登录 注册
找到 620 个结果

Data (in)equities in data science: Dissecting systemic and systematic biases in pulse oximetry

Data equity is an emerging framework for responsible data science. However, its core concepts, including fairness, representativeness, and information bias, remain largely abstract and general, lackin...

👤 Lillian Rountree | Harsh Parikh | Bhrama... 📰 arXiv 📅 2026 👁 56 📚 15

Distinguishability threshold for random geometric graphs

The spherical random geometric graph $G(n,d,p)$ is obtained by sampling $n$ independent points uniformly on the unit sphere $\mathbb{S}^{d-1}\subseteq\mathbb{R}^d$ and joining pairs of points which ar...

👤 Zach Hunter | Aleksa Milojević | Benny S... 📰 arXiv 📅 2026 👁 48 📚 15

Testing Centralized and Polycentric Computational Planning

This paper presents a reproducible synthetic benchmark comparing a computational planner, an agent-based market, and a hybrid meta-market within a common simulated economy. The benchmark incorporates ...

👤 Ricardo Alonzo Fernández Salguero 📰 arXiv 📅 2026 👁 46 📚 15

Calibrating Inelastic Markets to Options: The Lean Marketron and the Generalized Langevin Equation

The Marketron model of \cite{HalperinItkin2025Mark} and its option pricing extension in \cite{HalperinItkinMarketron2} suffer from structural non-identifiability: an eighteen-parameter space traps sol...

👤 Andrey Itkin 📰 arXiv 📅 2026 👁 39 📚 15

What If Consensus Lies? Selective-Complementary Reinforcement Learning at Test Time

Test-Time Reinforcement Learning (TTRL) enables Large Language Models (LLMs) to enhance reasoning capabilities on unlabeled test streams by deriving pseudo-rewards from majority voting consensus. Howe...

👤 Dong Yan|Jian Liang|Yanbo Wang|Shuo Lu|R... 📰 arXiv 📅 2026 👁 431 📚 14

Leveraging higher-order time integration methods for improved computational efficiency in a rainshaft model

Cloud and precipitation microphysics packages in atmospheric general circulation models typically use first-order time integration methods with a large time step, requiring ad hoc limiters and substep...

👤 Justin Dong|Sean P. Santos|Steven B. Rob... 📰 arXiv 📅 2026 👁 395 📚 14

VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning

Native visual reasoning treats visual generation as the medium of reasoning itself: visual states (i.e. images and videos) are not merely inputs to be understood or outputs to be rendered, but first-c...

👤 Junxiang Xu | Ruisi Wang | Fanyi Pu | Ma... 📰 arXiv 📅 2026 👁 214 📚 14

Certified Parallel-in-Time Sinkhorn for Dynamic Entropic Optimal Transport

Dynamic applications, including optimal-transport Flow Matching, repeatedly solve related entropic optimal transport problems, yet conventional distributed Sinkhorn processes frames sequentially and s...

👤 Xinyang Wen 📰 arXiv 📅 2026 👁 207 📚 14

Hindcast: Replaying Prediction Markets to Evaluate LLM Forecasters

Forecasters are evaluated by backtesting, which replays resolved questions and grades the probability the system would have assigned before the outcome was known. For LLMs, two channels leak the answe...

👤 Xiao Ye | Jacob Dineen | Evan Zhu | Shij... 📰 arXiv 📅 2026 👁 176 📚 14

The Matching Principle: A Geometric Theory of Loss Functions for Nuisance-Robust Representation Learning

Robustness, domain adaptation, photometric and occlusion invariance, compositional generalisation, temporal robustness, alignment safety, and classical anisotropic regularisation are usually treated a...

👤 Vishal Rajput 📰 arXiv 📅 2026 👁 164 📚 14

InterleaveThinker: Reinforcing Agentic Interleaved Generation

Recent image generators have demonstrated impressive photorealism and instruction-following capabilities in single-image generation and editing. However, constrained by their architectures, they canno...

👤 Dian Zheng | Harry Lee | Manyuan Zhang |... 📰 arXiv 📅 2026 👁 161 📚 14

What survives honest evaluation? Leakage-safe, search-aware assessment of LLM-driven trading strategy discovery

Large language models (LLMs) are increasingly used to discover trading strategies, and much of the resulting literature shares a methodological weakness: many candidate strategies are generated, the b...

👤 Eray Gençay 📰 arXiv 📅 2026 👁 122 📚 14

Learning structural balance of graphs from quantum spectral features

We develop a quantum approach to spectral feature extraction from the density of states (DOS) of a problem-dependent Hamiltonian, and apply it to machine learning on signed graphs. We propose to embed...

👤 Stefano Scali | Oleksandr Kyriienko 📰 arXiv 📅 2026 👁 111 📚 14

A Model for Imbalanced Label Aggregation: A Focus on Minority-Class Detection

We study imbalanced crowdsourcing with a focus on class-dependent annotator accuracy, a setting that, to the best of our knowledge, remains relatively underexplored despite its importance in real-worl...

👤 Gabriel Singer | Samuel Gruffaz | Olivie... 📰 arXiv 📅 2026 👁 107 📚 14

Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization

Pluralistic alignment has emerged as a critical frontier in the development of Large Language Models (LLMs), with reward models (RMs) serving as a central mechanism for capturing diverse human values....

👤 Qiyao Ma | Dechen Gao | Rui Cai | Boqi Z... 📰 arXiv 📅 2026 👁 106 📚 14

Estimating Flow Velocity and Vehicle Angle-of-Attack from Non-invasive Piezoelectric Structural Measurements Using Deep Learning

Accurate estimation of aerodynamic state variables such as freestream velocity and angle of attack (AoA) is important for aerodynamic load prediction, flight control, and model validation. This work p...

👤 Chandler B. Smith | S. Hales Swift | And... 📰 arXiv 📅 2026 👁 92 📚 14

Energy-Arena: A Dynamic Benchmark for Operational Energy Forecasting

Energy forecasting research faces a persistent comparability gap that makes it difficult to measure consistent progress over time. Reported accuracy gains are often not directly comparable because mod...

👤 Max Kleinebrahm | Jonathan Berrisch | Ph... 📰 arXiv 📅 2026 👁 81 📚 14

The Dual Nature of LLM Persona: Aggregated Tendencies and Frame-Dependent Geometry

Evaluations of LLM personas via psychometric questionnaires typically rely on aggregate scores, discarding within-instance correlation structure. We test whether this geometric structure is intrinsic ...

👤 Yuan Yuan 📰 arXiv 📅 2026 👁 80 📚 14

A Multiscale Ball Test for Conditional Mean Independence

Tests of conditional mean independence can lose power when departures are confined to a bounded part of a multivariate predictor space and the relevant spatial scale is unknown. We propose a Multiscal...

👤 Simon Rudkin | Wanling Rudkin 📰 arXiv 📅 2026 👁 63 📚 14

OccAny: Generalized Unconstrained Urban 3D Occupancy

Relying on in-domain annotations and precise sensor-rig priors, existing 3D occupancy prediction methods are limited in both scalability and out-of-domain generalization. While recent visual geometry ...

👤 Anh-Quan Cao | Tuan-Hung Vu 📰 arXiv 📅 2026 👁 60 📚 14
海洋智能体 🌊
海洋智能体
AI科研助手 · 3725篇文献
你在高级搜索页面,告诉我你想找什么方向的文献,我来帮你定位。