arXiv cs.LG - 2026-08-07 ​
260 items collected.
1. MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification ​
Author: Adam Simson, Ankush Dutta, Quang Bui
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2608.05196v1 Announce Type: new Abstract: Multiple sclerosis (MS) is diagnosed through clinical assessment, magnetic resonance imaging, laboratory evidence when appropriate, and exclusion of better explanations. Blood RNA expression data may contain disease associated immune signal, but a bloo...
2. When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters ​
Author: Fangxin Wang, Ziyi Zhang, Diyi Zhuang, Langzhou He, Shiyu Wang, Baichuan Mo, Philip S. Yu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05207v1 Announce Type: new Abstract: Frozen pretrained forecasters often fail in structured, recurring ways that are costly to repair through fine-tuning. We study corrective feature discovery: mining interpretable features of a frozen forecaster's residual to drive a lightweight post-hoc...
3. PPDL: LLM-Based Flows as Probabilistic Programs ​
Author: Louis Mandel, Guillaume Baudart, Mandana Vaziri, Martin Hirzel
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.PL
arXiv:2608.05234v1 Announce Type: new Abstract: Building reliable applications that leverage large language models (LLMs) remains a significant challenge. While LLMs offer impressive capabilities across diverse tasks, their outputs often lack accuracy and provide no clear measure of confidence. This...
4. Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language ​
Author: Xinran Feng, Yi Xie, Chao Zhang, Ruikun Li, Wanyun Ling, Ziyue Li, Chenxi Liu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05238v1 Announce Type: new Abstract: Training multimodal models to align time series with language runs into a self-supervision trap. The usual recipe asks an LLM to read a series and write a description, so label quality is capped by the perceptual skill the model is supposed to learn. T...
5. Disentangling 3D Modeling from Spatial Reasoning ​
Author: Haoze Sun, Jiequan Cui, Qingshan Xu, Richang Hong
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.05242v1 Announce Type: new Abstract: In this work, we explore an alternative paradigm for spatial reasoning by explicitly disentangling 3D perception from reasoning, rather than jointly acquiring implicit 3D perception and reasoning through large-scale training. Our key observation is tha...
6. Marginal Matching Does Not License Factorized Sampling: Auditing Conditional Style Leakage in Factorized Generative Models ​
Author: Duong Bach, Hai Nguyen Hong, Cuong Do
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.05243v1 Announce Type: new Abstract: Factorized generative models commonly regularize a latent style variable z_s by matching its marginal distribution to a fixed Gaussian prior and interpret this as evidence that the style representation is independent of class information. We show that ...
7. PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis ​
Author: Xiaomin He, Dongling Xiao, Jiahao Xie, Ruiqi Lu, Qianle Wang, Zhongbin Guo, Wanxuan Sun
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.05249v1 Announce Type: new Abstract: Real-world multimodal instructions often bundle multiple requirements with unequal importance, yet most multimodal training data still reduce instruction following to answering one self-contained question. We study this gap through \textbf{rubric compr...
8. Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning ​
Author: Yue Han, Ziniu Liu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05250v1 Announce Type: new Abstract: Multi-task supervised fine-tuning (SFT) often casts a heterogeneous data mixture as a single optimization problem, even though different tasks may reach their best generalization at different times. msft exposes this mismatch through task-wise roll-out...
9. Beyond Rotations: AuroOFT for Expressive Quantized Orthogonal Fine-Tuning ​
Author: Yue Han, Dianlin Wang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05253v1 Announce Type: new Abstract: Quantized orthogonal fine-tuning (qoft) enables parameter-efficient adaptation of low-bit language models by learning structured activation rotations before frozen quantized weights. However, its task-specific updates remain constrained to linear ortho...
10. An Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goals ​
Author: Ramin Pishehvar
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2608.05255v1 Announce Type: new Abstract: Retail investors lack access to the kind of personalized, tax-aware portfolio management that institutional clients take for granted -- existing robo-advisors use static, rule-based allocation, and institutional-grade systems require account minimums a...
11. Evaluating Machine Learning Models for Post-Wildfire Debris-Flow Prediction ​
Author: Quinn Ledingham, Zhengsen Xu, Yimin Zhu, Zack Dewis, Mabel Heffring, Saeid Taleghanidoozdoozan, Motasem Alkayid, Megan Greenwood, Lincoln Linlin Xu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.05265v1 Announce Type: new Abstract: Prediction of post-wildfire debris flows is critical for mitigating hazards to communities, infrastructure, and resources during intense rainfall in recently burned areas. However, identifying reliable machine learning models is complicated by overlapp...
12. Rectifying Geometric Misalignment: Online Source-Free Adaptation for Class-Imbalanced EEG ​
Author: Shiwen Chu, Shanglin Li, Motoaki Kawanabe, Reinmar Kobler
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05315v1 Announce Type: new Abstract: Electroencephalography (EEG) based Brain-Computer Interfaces (BCIs) often require unsupervised domain adaptation (UDA) to generalize across subjects and sessions. While Riemannian alignment methods like the Riemannian Centering Transformation (RCT) are...
13. QEvict: Recoverable Quantized KV Eviction for Attention-Drift-Robust Long-Context Decoding ​
Author: Ayushman Garg, Akshita Gupta, Shaswata Bhattacharya, Abhishek Gupta, Sandeep Kumar, Manoj Kumar
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.05326v1 Announce Type: new Abstract: Autoregressive large language model inference is increasingly constrained by the memory footprint of the Key-Value (KV) cache. A dominant line of work reduces this footprint by evicting tokens that appear unimportant under attention-derived scores. How...
14. DG-FedReuse: Proxy-Gradient-Gated Cached-Update Reuse with Matched Sparse Uplink Accounting ​
Author: Rahil Aftab, Vineet Kumar Rakesh, Soumya Mazumdar, Tapas Samanta
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2608.05358v1 Announce Type: new Abstract: Federated learning repeatedly incurs local optimization and model-update transmission. We study DG-FedReuse, a simulator-level mechanism that allows selected clients to contribute age-decayed cached updates when a stochastic head-gradient discrepancy p...
15. Quantum-Structured World Models (QSWMs) for Predictive Latent Dynamics ​
Author: Hailong Jiang, Emran Hossain, Feng Yu, Jianfeng Zhu, Guilin Zhang, Wulan Guo
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05371v1 Announce Type: new Abstract: World models learn latent states that summarize interaction histories, evolve over time, and support prediction, simulation, or planning. Most existing world models represent these states using classical vectors, probability distributions, recurrent hi...
16. Spectral Distillation: From Nonlinear Dynamics to Linear State-Space Models ​
Author: Liane Galanti, Devan Shah, Shlomo Fortgang, Elad Hazan
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05416v1 Announce Type: new Abstract: Can nonlinear dynamical systems be learned through a compact linear state-space representation, without directly solving a non-convex system-identification problem? We give a provable pipeline for doing so. Starting from observations of an unknown nonl...
17. Perturbation Sensitivity at Convergence: A Simple Signal for Identifying Spuriously Correlated Samples ​
Author: Nilesh Kumar
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.05419v1 Announce Type: new Abstract: Models trained by empirical risk minimization on data containing spurious correlations achieve high average accuracy while failing on subpopulations where the correlation does not hold. Existing methods for identifying the affected samples without grou...
18. IFlowNets: Extending Generative Samplers to Learn Strategies in Incomplete Information Games ​
Author: Conor M. Artman, Nicholas Di, Scott Perkins
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.MA
arXiv:2608.05422v1 Announce Type: new Abstract: While many algorithms blend reinforcement learning (RL) with counterfactual regret (CFR) methods to leverage tradeoffs in computational speed and performance, there are fewer investigations into generative sampling frameworks in game theoretic applicat...
19. Why the Third Axis Is Freedom ​
Author: Michael Timothy Bennett
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.05423v1 Announce Type: new Abstract: In generative training, a model produces an output and is penalised for its difference from an example. With one output per comparison, a model that produces one common answer can outperform a model retaining a broader repertoire. Explorative Modeling ...
20. EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents ​
Author: Xuying Ning, Dongqi Fu, Tianxin Wei, Hanqing Zeng, Yuanchen Bei, Bingxuan Li, Zihao Li, Qifan Wang, Xiang Shen, Yifan Wu, Jiayi Liu, Hong Li, Yinglong Xia, Xiangjun Fan, Hanghang Tong, Jingrui He
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.05446v1 Announce Type: new Abstract: Long-horizon LLM agents increasingly rely on external execution support to maintain state, track progress, invoke tools, verify outcomes, and reuse experience across interactions. However, effective harness use raises two coupled challenges: state form...
21. Hybrid Probabilistic Zonotopes for Identifiable and Refinable Predictive Uncertainty ​
Author: Zhen Zhang, Amr Alanwar
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.05454v1 Announce Type: new Abstract: Probabilistic prediction heads in neural networks typically output either a Gaussian mixture or a single conformal region. Neither separates the distinct sources of uncertainty often present in real prediction tasks: a discrete choice among modes, boun...
22. Matrix Zonotopic Attention: A Context-Adaptive Value Projection for Set Transformers ​
Author: Zhen Zhang, Amr Alanwar
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.05472v1 Announce Type: new Abstract: Multi-head attention combines an input-dependent softmax routing with an input-independent linear value projection, so the per-sample operator mapping aggregated values to outputs is the same for every input set. We study the consequences of this asymm...
23. KV-Skill: Forging Expertise in the Model's Native Language ​
Author: Zhaowei Han, Xiang Zhang, Bing Han, Kai Liu, Danqi Hu, Jie Liu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05475v1 Announce Type: new Abstract: Task knowledge is commonly stored either as text in the prompt or as an update to model weights. Text is modular but must be interpreted on every use, while weight adaptation makes the resulting capability difficult to load, remove, or share independen...
24. Align-RAG: Alignment Is All You Need for TSFM In-Context Learning ​
Author: Mohammad Asadi, Soheil Hor, Bardiya Akhbari, Jack W. O'Sullivan, Tahoura Nedaee, Layne C. Price, Raviteja Anantha, Euan Ashley, Ehsan Adeli
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2608.05571v1 Announce Type: new Abstract: Retrieval-augmented forecasting promises to adapt frozen Time Series Foundation Models (TSFMs) to new domains without fine-tuning, but recent methods typically rely on learned fusion modules, i.e., trained adapters that merge retrieved examples into th...
25. LC-GRPO: Bridging Train-Inference Gap for Flow-Based GRPO with Langevin Correction ​
Author: Yingqing Guo, Hui Yuan, Zijian He, Mengdi Wang, Zheng Ding
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.05600v1 Announce Type: new Abstract: Flow-based generative models are typically sampled by solving a deterministic ordinary differential equation (ODE), whereas online reinforcement learning requires stochastic rollouts for policy exploration and optimization. Existing GRPO methods for fl...
26. GAUGE: Granularity-Adaptive Counterfactual Gating of Evidence for Incomplete Multimodal Classification ​
Author: Yunping Shi, En Yu, Kairui Guo, Jie Lu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.05608v1 Announce Type: new Abstract: Multimodal classification typically assumes all modalities are available, yet real-world inputs are often incomplete. Imputation and dynamic fusion can mitigate such incompleteness, but existing methods operate at a coarse modality level and thus canno...
27. Reasoning Errors Have a Region and a Direction in the Residual-Stream Trajectory of LLMs ​
Author: Hamed Damirchi, Ignacio Meza De la Jara, Damith Ranasinghe, Yuhang Liu, Javen Shi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.05660v1 Announce Type: new Abstract: As language models are increasingly used for tasks that require verifiable reasoning, reliably distinguishing sound reasoning from flawed reasoning has become an important practical problem. Recent trajectory-based methods seek this signal in layerwise...
28. Potential Matching Optimal Transport: Continuous Normalizing Flows for Exact $p$-Wasserstein Dynamics ​
Author: Lishuo Zhang (School of Mathematical Sciences, Shanghai Jiao Tong University), Ruizhi Huang (School of Mathematical Sciences, Shanghai Jiao Tong University), Yang Yu (School of Mathematical Sciences, Shanghai Jiao Tong University), Lei Li (School of Mathematical Sciences, Shanghai Jiao Tong University, Institute of Natural Sciences, MOE-LSC, Shanghai Jiao Tong University)
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.05666v1 Announce Type: new Abstract: We introduce Potential Matching Optimal Transport (PMOT), a potential-flow framework for general $p$-cost optimal transport with $c_p(x,y)=|x-y|^p$. PMOT parameterizes the CNF velocity field with a scalar potential in the generalized Benamou--Brenier...
29. When Does Consensus Mean Correctness? Measuring the Agreement-Accuracy Coupling with Semantics-Preserving Re-Rendering ​
Author: Rasul Khanbayov, Hasan Kurban
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05670v1 Announce Type: new Abstract: A model's agreement across perturbed inputs is used both as a label-free reliability signal and as a self-training target, on the premise that agreement tracks correctness. That coupling is rarely measured directly: natural-image perturbations preserve...
30. Consistency Has a Computable Blind Spot: A Commutation Theory of Label-Free Reliability for Vision-Language Figure Reading ​
Author: Rasul Khanbayov, Hasan Kurban
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05675v1 Announce Type: new Abstract: Label-free reliability for vision-language models rests on invariance: perturb the input and a faithful reader's answer should not change. This has a known blind spot, a systematic misreading survives the perturbation and gets certified wrong, which we...
31. SEAM: Global consistency beyond local accuracy in scientific machine learning ​
Author: Gnankan Landry Regis N'guessan, Bum Jun Kim
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2608.05702v1 Announce Type: new Abstract: Scientific machine learning commonly validates models at the level of a subdomain, a benchmark split, or an explanation for one prediction. Yet such local checks cannot establish whether the resulting explanations can be assembled into one globally adm...
32. Spectral Aliasing Pretext: A novel task for Self-Supervised fault diagnosis in rotating machinery ​
Author: Victor Gialis, Maxime Metz, David Esteve, Abdenour Soualhi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.05705v1 Announce Type: new Abstract: Deep learning is a new way for machinery fault diagnosis but requires extensive labeled data, a scarce resource in industrial settings. We propose Spectral Aliasing Pretext (SAP), a self-supervised learning method that pretrains models on unlabeled vib...
33. CircuitSteer: Geometrically Aligned Multi-Layer Steering via Sparse Autoencoder Circuits ​
Author: Mehrshad Saadatinia, Parsa Razmara, Ardalan Aryashad, Ali Abbasi, Seyedarmin Azizi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05732v1 Announce Type: new Abstract: Controlling the behavior of large language models (LLMs) remains a critical challenge for AI alignment. Existing steering methods, such as Contrastive Activation Addition (CAA), typically rely on fixed single-layer interventions derived from aggregate ...
34. Multivariate Time Series Forecasting needs Cross Variable Loss ​
Author: Kuiye Ding, Yifan Hu, Hanchen Wang, Hao Xue
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.05742v1 Announce Type: new Abstract: Multivariate time series forecasting presents unique challenges because future variables often co-evolve under shared system dynamics. While existing studies mainly focus on cross-variable dependencies in historical observations, dependencies among fut...
35. Equipment-centric workpiece localization in near real-time using deep learning-based vision and event-driven finite state machines ​
Author: Dohyeon Kong, Jaebong Cho, Hyunbo Cho
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05744v1 Announce Type: new Abstract: Continuous workpiece localization is essential for traceability and process coordination in hot forging, but direct tracking is unreliable because of extreme temperatures, surface degradation, and irregular routing. This study presents an equipment-cen...
36. Accelerating nanodrug development in continuous flow systems using informed prediction models based on low-cost surrogate nanoparticles ​
Author: Kai Dahms, Eilien Heinrich, Jochen Schmid, Michael Bortz, Iryna Savych, Regina Bleul
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2608.05761v1 Announce Type: new Abstract: The development of nanotherapeutics often involves extensive empirical optimization due to the sensitivity of nanoparticle properties, such as size and polydispersity index (PDI), to minor changes in process parameters. Factors like formulation concent...
37. Neuro-Symbolic Closed-Loop Control of Laser Powder Bed Fusion with an In-Loop Ontology ​
Author: Gisuk Hong, Jaebong Cho, Hyunbo Cho
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05773v1 Announce Type: new Abstract: A geometry-conditioned, neuro-symbolic closed-loop architecture is proposed for laser powder bed fusion, in which a standards-aligned ontology operates inside the control loop and couples symbolic reasoning with statistical learning to set the targets ...
38. GROM: Gradient-Free Rapid One-Shot Machine Unlearning ​
Author: Pawe{\l} Batorski, Przemys{\l}aw Spurek, Paul Swoboda
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.05783v1 Announce Type: new Abstract: Machine unlearning has become a critical capability for safely removing specific, sensitive knowledge from large language models (LLMs). Current state-of-the-art approaches primarily rely on iterative, training-time unlearning via fine-tuning. However,...
39. Predicting Task Difficulty Without Rollouts ​
Author: Stefan Krsteski, Charlotte Meyer
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.05797v1 Announce Type: new Abstract: Task difficulty dictates an agent's likelihood of success, and estimating it without rollouts means forecasting this directly from a task description before executing costly simulations in stateful environments. Reliable estimates would therefore allow...
40. Learning to Rank Tensor Network Contraction Plans for GPU-Accelerated Quantum Circuit Simulation ​
Author: Alfred M. Pastor, Maribel Castillo, Jose M. Badia
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.PF, quant-ph
arXiv:2608.05819v1 Announce Type: new Abstract: Classical simulation remains essential for developing and validating quantum algorithms, but its cost grows rapidly with circuit size. Tensor-network contraction can reduce this cost by exploiting circuit structure, although its efficiency depends stro...
41. Evidential Rule Learning for Interpretable Classification with Abstention ​
Author: Javier Fumanal-Idocin, Javier Andreu-Perez
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.05859v1 Announce Type: new Abstract: Interpretable classification often requires more than accurate predictions for real-life deployment: models should be transparent about the evidence behind their decisions and abstain when they cannot decide reliably. We introduce Fast Evidential Rule ...
42. Beyond Feature Importance: A Comparative Analysis of Pattern Detection Methods in Cluster Interpretation ​
Author: Benjamin Connor, Anna Jurek-Loughrey, Lu Bai, Muhammad Fahim
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.05880v1 Announce Type: new Abstract: Interpreting clustering outcomes remains a fundamental challenge in data analysis, particularly in domains such as healthcare where meaningful patterns must be extracted from high-dimensional data. While numerous explainability techniques exist, they a...
43. Alternating Levenberg-Marquardt Training of Physics-Informed Neural Networks with Fourier-Enhanced Features ​
Author: Yulun Wu, Matthieu Barreau, Miguel Aguiar, Karl H. Johansson
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.05892v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) often fail to accurately resolve partial differential equations (PDEs) with high-frequency or multi-scale solutions, as well as strongly nonlinear problems. Two factors underlie this difficulty: spectral bias, t...
44. CohortHijack: Robustness of Single Cell Annotation to Companion Cell Removal ​
Author: Arash Vashagh, Yasmin Vashagh
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05900v1 Announce Type: new Abstract: Many single-cell annotation tools refine an initial cell label using nearby cells or cluster-level voting. We study whether this refinement can be manipulated without changing the target cell. We introduce CohortHijack, a robustness audit that removes ...
45. BioM-JEPA: joint-embedding prediction of graph-connected gene blocks in single cells ​
Author: Yuhao Wang, Zelin Zang, Yuxuan Liu, Zhen Lei, Stan Z. Li
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05928v1 Announce Type: new Abstract: Single-cell transcriptomes are sparse observations of coordinated biological programmes, yet most self-supervised models learn by reconstructing individual genes. Here we present BioM-JEPA, a joint-embedding predictive architecture that instead predict...
46. How Far Do Simple Transformations Translate Across Text Embedding Models? ​
Author: Sid Ali Hamideche (Orange Research), Louis Adrien Dufrene (Orange Research), Quentin Lampin (Orange Research), Guillaume Larue (Orange Research)
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05980v1 Announce Type: new Abstract: We investigate whether simple transformations can translate representations across heterogeneous text embedding models. Understanding how independently trained models organize semantic information is an enabler for AI-to-AI latent communication without...
47. THBKG: A Temporal Biomedical Knowledge Graph for Decision-Aligned Clinical Advancement Prediction ​
Author: Pui Chung Siu, Claudia Cabrera, Mani Mudaliar, Arkaitz Zubiaga
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2608.05982v1 Announce Type: new Abstract: Inadequate target--disease linkage accounts for 40--50% of Phase~II efficacy failures, so anticipating which programmes will advance would let sponsors back the hypotheses most likely to reach patients. What a programme can be judged on is the evidenc...
48. Observation-Grounded Self-Predictive Reinforcement Learning for Visual Continuous Control ​
Author: Xinwei Liu, Junyuan Liang, Jianting Zhang, Wuhui Chen
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2608.05989v1 Announce Type: new Abstract: Sample-efficient policy learning from pixels is a long-standing challenge in reinforcement learning (RL). Recent dynamics-based representation learning methods have significantly improved the sample efficiency of model-free visual RL by learning dynami...
49. A Unified Risk View of Uncertainty: Posterior Risk for Disentanglement and Evaluation Beyond Proxies ​
Author: Frieder Wizgall, Georg Tirpitz, Moritz Seiler, Kerstin Ritter, B'alint Mucs'anyi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.05995v1 Announce Type: new Abstract: Reliable uncertainty estimates are critical in safety-sensitive applications, where understanding the sources of predictive uncertainty is essential. This often requires disentangling epistemic uncertainty from aleatoric uncertainty, yet these uncertai...
50. Do Tabular Foundation Models Agree with Themselves? ​
Author: Christian Kl"otergens, Vijaya Krishna Yalavarthi, Lars Schmidt-Thieme, Tom Hanika
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06004v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) are currently the best approach to tabular prediction problems. They are constructed as transformers that approximate the Bayesian posterior predictive distribution based on a pre-training prior. These univariate predic...
51. ProDVI: Programmatic Dynamics Priors for Value Network Initialization ​
Author: Xinwei Liu, Junyuan Liang, Jianting Zhang, Wuhui Chen
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.06015v1 Announce Type: new Abstract: Deep Reinforcement Learning (RL) is notoriously sample inefficient. One contributing factor is that RL agents are typically initialized from scratch, forcing them to acquire task-relevant knowledge through online interaction. Existing approaches obtain...
52. BioKD: Selective Physiology-to-Video Knowledge Distillation via Reliability Gate for Emotion Recognition ​
Author: Bojing Hou, Ruohao Li, Yitong Zhu, Hongjun Liu, Luwen Yu, Yuyang Wang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06023v1 Announce Type: new Abstract: To address the limitations of video-based emotion recognition under ambiguous or socially masked behavioral cues, as well as the poor deployability of physiological signals, this paper proposes a reliability-aware physiology-to-video knowledge distilla...
53. Hybrid-Adaptive Thread Tuning to Mitigate Simulation Execution Bottlenecks in High-Performance Reinforcement Learning Inference ​
Author: Jiming Su, Hantao Hua, Lujia Yin, Yiping Yao, Feng Zhu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.MA, cs.PF
arXiv:2608.06025v1 Announce Type: new Abstract: In simulation-in-the-loop decision-making systems, reinforcement learning (RL) inference is often constrained by simulator-side execution overhead, where workloads are highly dynamic and sensitive to runtime thread configurations. Existing multithreade...
54. Dynamic Graph Prompting via Topology-Routed Mixed-Curvature Experts ​
Author: Quanxin Wang, Xuanting Xie, Bingheng Li, Xingtong Yu, Shuo Wang, Ruiyi Fang, Zhao Kang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06031v1 Announce Type: new Abstract: Dynamic graph prompting freezes a pre-trained temporal backbone and adapts it to label-scarce downstream tasks using lightweight prompts. However, existing methods operate within a single, fixed embedding space. In this work, we reveal that temporal sh...
55. Does Latent Context Help? A Controlled Evaluation of Inverse Reinforcement Learning in Arctic Shipping ​
Author: Vaishnav Vaidheeswaran, Dilith Jayakody, Biruk Ambaw, Jaswanth Kumar, Md Mahbub Alam, Gabriel Spadon
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA
arXiv:2608.06105v1 Announce Type: new Abstract: Artificial Intelligence (AI)-assisted navigation can help Arctic shipping adapt to rapidly changing sea-ice conditions, but reliable deployment requires reward models that are interpretable and robust to changing environments. Inverse reinforcement lea...
56. Kastor: An efficient fine-tuning strategy for generative emulation of PDE simulations ​
Author: Guillaume Couairon, Alexis Jacq, Yu-Han Wu, Renu Singh, Yana Hasson, Quentin Berthet, Romuald Elie
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.06107v1 Announce Type: new Abstract: Machine learning offers a promising avenue to accelerate physical simulations by replacing computationally expensive traditional Partial Differential Equation (PDE) solvers with fast, differentiable surrogate models. However, standard auto-regressive M...
57. Is Self-Pretraining really useful to improve diagnosis in medical Time Series? ​
Author: Omar Coser, Antonio Orvieto, Paolo Soda, Loredana Zollo
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.06122v1 Announce Type: new Abstract: Inspired by recent evidence that transformer architectures benefit from Self-PreTraining (SPT) on long-context benchmarks, we investigate whether similar gains extend to multimodal, multivariate, and even simple univariate medical time series. Our obje...
58. LLM Inference Under Bursty Workload Distribution: Modifying the WAIT Algorithm ​
Author: Anjali Gangadhar Katageria, Shobha Rani, Raghu Nandan Sengupta
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06135v1 Announce Type: new Abstract: Large Language Models (LLMs) such as ChatGPT and Claude are widely used for information retrieval and problem-solving. Recent work has focused on improving scheduling algorithms to boost throughput while maintaining low latency. However, these approach...
59. SkillTFM: Gated Skill Evolution for Training-Free Adaptation of Tabular Foundation Models ​
Author: Yi He, Zhengkang Guan, Anpeng Wu, Peng Cui, Fei Wu, Kun Kuang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06137v1 Announce Type: new Abstract: Tabular data are ubiquitous in real-world applications and are crucial for data-driven prediction and decision-making across science, industry, finance, healthcare, and public services. Tabular foundation models (TFMs) have emerged as a promising parad...
60. Threshold-Based Early Stopping of Accumulations in Neural Networks with Binary Activation ​
Author: Quentin Luquet de Saint-Germain, Massil Ait Abdeslam, Jean Pierre David
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2608.06177v1 Announce Type: new Abstract: Binary neural networks are very attractive for constrained deployment, enabling small footprint and low-power inference. For binary activations, the dot products become sign-controlled additions or subtractions, but the number of operations is unchange...
61. SAGA: Score-Weighted Adaptive Generation Alignment for Low-Resource Nordic Language Models ​
Author: Hoda Fakharzadehjahromy, Emil Wiman, Andreas Bueff, Hafsteinn Einarsson, Fredrik Heintz
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06179v1 Announce Type: new Abstract: Preference optimisation has proven effective for improving large language models but typically relies on costly human preference annotations. Extending these methods to morphologically rich, low-resource languages remains challenging because such annot...
62. Continual Learning in Transition ​
Author: Zhiyan Hou, Dan Zhang, Tao Feng, Liyuan Wang, Wei Li, Xiangzhao Hao, Hongyan An, Junfeng Fang, Haokai Ma, Zhaohui Xu, Haiyun Guo, Jinqiao Wang, Tat-Seng Chua
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.06216v1 Announce Type: new Abstract: Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g., training strategies, architectural designs, and weight adaptation. However, emerging paradigms are res...
63. Timestep-Conditioned Transformers for Global Weather Forecasting ​
Author: Sam Levang, Fran Bartolic, Ty Dickinson, Chase Dwelle, Paulius Rauba, Viktor Cikojevic
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.OS
arXiv:2608.06241v1 Announce Type: new Abstract: Existing machine-learning weather forecasting models rely on predetermined and fixed autoregressive timesteps. The choice of model timestep involves a fundamental trade-off: shorter timesteps (e.g. 1 to 6 hours) finely resolve atmospheric dynamics with...
64. A Six-Dimensional Taxonomy of Post-Training Adaptation Techniques with Applications in AI Governance ​
Author: Fardin Afdideh, Fernando Seoane, Farhad Abtahi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06246v1 Announce Type: new Abstract: Post-training adaptation has become central to modern machine learning practice and includes techniques such as retraining, fine-tuning, parameter-efficient adaptation, alignment, retrieval augmentation, model editing, unlearning, calibration, and Mult...
65. MetaboLLM: a metabolomics-specialized large language model for biochemical knowledge integration and predictive metabolite graph construction ​
Author: Dohyun Ku, Min Gu Kwak, Francisco J. Pasquel, Jing Li
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06253v1 Announce Type: new Abstract: Metabolomics knowledge is distributed across heterogeneous resources and remains difficult to translate into predictive representations. We developed MetaboLLM, a metabolomics-specialized large language model adapted through continual pretraining, supe...
66. RxnCLF: Contrastive Transformation-Aware Reaction Foundation Model for Improved Reactivity Prediction ​
Author: Yiting Zheng, Cheng Fang, Anthony Donofrio, Haote Li
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06259v1 Announce Type: new Abstract: Reaction yield prediction remains challenging because labeled data are scarce and reaction space is both combinatorially large and sparsely populated, limiting the generalization of existing reaction representations. String-, fingerprint-, and graph-ba...
67. Hypothesis Testing with Conditional Queries: Learnability and the Value of Interaction ​
Author: Zonghuan Xu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.TH
arXiv:2608.06262v1 Announce Type: new Abstract: Model evaluations may fix all tests before observing any responses or select later tests using earlier responses. We study this choice in a conditional-query model on a finite outcome space $\mathcal{X}$ with $|\mathcal{X}|=N$. We first ask which pairs...
68. The Tamed Subgradient Unadjusted Langevin Algorithm beyond Convexity ​
Author: Iosif Lytras, Nikolaos Makras, Sotirios Sabanis
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, math.OC, math.PR, stat.ML
arXiv:2608.06283v1 Announce Type: new Abstract: We study the problem of sampling from target distributions whose potentials are simultaneously non-smooth, subject to superlinear gradient growth, and non-convex. We introduce the Subgradient Tamed Unadjusted Langevin Algorithm (SG-TULA), a discretisat...
69. Surv-IPTB: An Attention-Based Model for Estimating Individual Probability of Treatment Benefit with Survival Data ​
Author: Lev V. Utkin, Stanislav K. Kogan, Andrei V. Konstantinov
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.06288v1 Announce Type: new Abstract: This work presents a novel attention-based framework for estimating the Individual Probability of Treatment Benefit (IPTB) in survival analysis contexts. The proposed model, called Surv-IPTB, directly quantifies the probability that a specific patient ...
70. BaKron: Efficient Quantization with Kronecker-Factored Hessians ​
Author: Johann Birnick, Rayan Saab
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.06291v1 Announce Type: new Abstract: We accelerate a family of algorithms for neural network quantization whose geometry is informed by any Kronecker-factored approximation of the Hessian. GPTQ-style adaptive rounding typically uses one-sided information derived from input activations. Tw...
71. On-Policy Self-Distillation without Any Supervision ​
Author: Yijiang Li, Bingyang Wang, Yijun Liang, Yunjie Tian, Di Fu, Nuno Vasconcelos
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06296v1 Announce Type: new Abstract: On-policy (Self-)Distillation (OPD / OPSD) has shown strong potential for post-training large language models (LLMs). However, existing methods still rely heavily on external supervision, including ground-truth signals, environmental feedback, or guida...
72. RRC: Unlocking Generative Reward Models in LLM Reinforcement Learning via Ranking-Based Reward Construction ​
Author: Chenglong Wang, Ziming Zhu, Yifu Huo, Bei Li, Qiaozhi He, Yan Ding, Xiaoyang Hao, Yuxin Gao, Tianhua Zhou, Xiaojia Chang, Tongran Liu, Jingbo Zhu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.06310v1 Announce Type: new Abstract: Recent advances in reward modeling show a paradigm shift from discriminative reward models to generative reward models. However, despite their strong capabilities in response ranking, generative reward models have not realized their potential in reinfo...
73. CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks ​
Author: Fanzhe Meng, Guoxin Chen, Jiale Zhao, Shuang Sun, Zhiyu Lin, Wayne Xin Zhao, Ruihua Song, Ji-Rong Wen, Kai Jia
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.06352v1 Announce Type: new Abstract: Training terminal agents requires executable and verifiable tasks that are not merely solvable, but appropriately challenging for learning. Executable validation establishes feasibility, yet does not reveal how a task behaves relative to a given solver...
74. An Optimal Agnostic PAC Algorithm ​
Author: Markus Engelund Mathiasen, Jian Qian, Nikita Zhivotovskiy
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DS, math.ST, stat.TH
arXiv:2608.06363v1 Announce Type: new Abstract: Let $H\subseteq{-1,+1}^X$ be a class of finite VC dimension $d\ge1$. Writing $L$ for the binary risk and $L^*=\min_{h\in H}L(h)$, we construct a learner achieving the statistically optimal risk bound: from an i.i.d.\ sample of size $n$, for every $0<...
75. Safe Evolution with Circuit Anchors ​
Author: Yan Liu, Jie Fu, Tsung-Yi Ho
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.NE
arXiv:2608.05158v1 Announce Type: cross Abstract: In biological evolution, unconstrained mutation can lead to catastrophic outcomes: organisms may evolve enhanced capabilities while losing essential functions for survival. Nature's solution is \textit{developmental constraints}, where core regulator...
76. The Ignition Index: Measuring Global Workspace Dynamics in Language Models ​
Author: Saman Rahbar
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.05160v1 Announce Type: cross Abstract: We introduce the Ignition Index (I), a validated scalar metric that operationalizes Global Workspace Theory's (GWT) all-or-none ignition prediction in transformer language models. The metric fits a four-parameter sigmoid to per-layer linear probe acc...
77. SemiAdapt-Instruct: Extensible Instruction Tuning via Latent Domain-Specialised Adapters ​
Author: Josh McGiff, Salma Mekaoui, Robert Shanahan, Nikola S. Nikolov
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.05161v1 Announce Type: cross Abstract: Instruction-tuned LLMs are deployed into environments where domains evolve, yet extending a fine-tuned model's capabilities without full retraining remains an unsolved practical challenge. We present SemiAdapt-Instruct, a modular framework that disco...
78. PoolBench: A Benchmark for Pooling Strategies in Concept Representation Evaluation for Decoder-Only LLMs ​
Author: Ayushi Agarwal
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.05162v1 Announce Type: cross Abstract: Pooling is a consequential but under-examined design choice in decoder-only concept representation work: practitioners must collapse token-level hidden states into a passage-level vector, yet no shared protocol exists for comparing this choice across...
79. Where Privacy Risk Lives in English-Source Multilingual RAG: A Stage-Decomposed Audit Across Five Query Languages ​
Author: Yanhang Li, Zhichao Fan, Zexin Zhuang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.05163v1 Announce Type: cross Abstract: A common assumption holds that switching to a non-English language makes a multilingual RAG system easier to attack for personal information. We test this on an English-source synthetic-PII corpus with five query languages and a two-stage defence (LL...
80. Cross-Architecture Steering Transfer in Language Models: A Systematic Empirical Study ​
Author: Ayushi Agarwal
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.05164v1 Announce Type: cross Abstract: Independently trained large language models may develop shared internal representations of semantic concepts despite architectural differences -- but whether this geometric similarity has functional consequences for cross-model behavioural control re...
81. A Study of ASR Adaptation and Representation Dimensionality Reduction in Persian Speech Emotion Recognition Using Whisper ​
Author: Ali Shendabadi, Parnia Izadirad, Mostafa Salehi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.SD
arXiv:2608.05165v1 Announce Type: cross Abstract: Speech Emotion Recognition (SER) in low-resource languages remains a challenging problem due to limited labeled data. In this work, we study the use of Whisper for Persian SER with a particular focus on representation dimensionality reduction and lan...
82. Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning ​
Author: Sachini Weerasekara, Sagar Kamarthi, Jacqueline Isaacs
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.CY, cs.HC, cs.LG
arXiv:2608.05166v1 Announce Type: cross Abstract: We present an evaluation of cognitive bias expression in state-of-the-art instruction-tuned LLMs under realistic multi-turn interaction settings. Our work introduces a novel three-condition experimental framework that disentangles the effect of expos...
83. Challenges for Musical Education in the Age of AI and Digital Transformation ​
Author: Jean-Pierre Briot
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG
arXiv:2608.05176v1 Announce Type: cross Abstract: Music education has never been a static discipline. Each major technological shift has forced educators and institutions to reconsider what they teach, how they teach it, and why. We now stand at what may be the most consequential of such turning poi...
84. CLARA: Clarification of Language Ambiguity through Result Analysis for Natural-Language Cancer Genomics Queries ​
Author: Pratyush Kumar Shukla, Manveer Singh Tib, Siddhant Garg
Published: 8/7/2026, 4:00:00 AM
Categories: q-bio.GN, cs.LG, q-bio.QM
arXiv:2608.05195v1 Announce Type: cross Abstract: A natural language interface can be used to make cancer genomics databases easier to use, but even if a question is perfectly fluent, its scientific meaning can be ambiguous. We propose CLARA, a framework that represents a question as a typed scienti...
85. From Continuous Predictors to Clinical Thresholds: Early Evidence on Performance Trade-offs of Guideline-Based Categorisation for Ischaemic Stroke Outcome Prediction ​
Author: Esra Zihni, Katryna Cisek, Hamzah Ziadeh, Hendrik Knoche, Robert Mikulik, John D. Kelleher
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.05203v1 Announce Type: cross Abstract: Machine learning models achieve strong predictive accuracy for 90-day outcome prediction in acute ischaemic stroke, yet clinical adoption is limited by the misalignment of model explanations with clinicians' reasoning. Motivated by a clinician user s...
86. SkillTrace: Multi-Trace Provenance Auditing for LLM-Agent Skill Reuse ​
Author: Jialuo Chen, Minghe Wang, Lingqi Jiang, Jianan Ma, Xinhao Deng, Xiaohu Du, Ruixiao Lin, Yunhao Feng, Linkang Du, Jingyi Wang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.05204v1 Announce Type: cross Abstract: LLM-agent ecosystems are rapidly growing around reusable skills: mixed-modality packages of metadata, natural-language instructions, code, tools, references, and operational workflows. As skills become marketplace artifacts, auditing their reuse is n...
87. Otter: A Time-Aware, History-Conditioned Human Chess AI ​
Author: Tarun Kumar S
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.05206v1 Announce Type: cross Abstract: Otter is a 15.3M-parameter human chess AI that predicts human move selection by modeling play as a time-aware, sequential process rather than treating each position in isolation. It combines two conditioning signals: (1) a move history encoder that c...
88. Quality Diversity for Reliable Data Driven Time-Use Optimization ​
Author: Aneta Neumann, Ty Stanford, Dorothea Dumuid, Frank Neumann
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, cs.NE
arXiv:2608.05230v1 Announce Type: cross Abstract: The daily allocation of the finite 24-hour time budget is strongly associated with physical, mental, and cognitive health. While predictive models can estimate the relationship between time-use compositions and health outcomes such as body mass index...
89. Analysis of Numerical Localisation in LLM Translations ​
Author: Patrizia Kaye
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.05232v1 Announce Type: cross Abstract: The work of Tang et. al. (2025) on numerical translation is extended by analysing the capability of five large language models (LLMs) for the localisation of times, numbers, and dates instead of translation. Models were selected that could be loaded ...
90. One Qubit Can Beat One Bit: Quantum Advantage for Post-Training Quantization ​
Author: Yuma Ichikawa, Moeto Mishima
Published: 8/7/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG
arXiv:2608.05240v1 Announce Type: cross Abstract: One-bit post-training quantization represents each weight using only its sign, requiring all deployment contexts to share the same binary weight matrix even when their activation statistics favor different sign patterns. We study this shared-sign con...
91. A Unified Causal Inference Framework for the Desirability of Outcome Ranking Paradigm in Benefit-Risk Evaluation ​
Author: Yuan Feng, Shiyu Shu, Yixin Fang, Ionut Bebu, Toshimitsu Hamasaki, Scott Evans, Guoqing Diao
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2608.05244v1 Announce Type: cross Abstract: We developed a unified covariate-adjusted causal inference framework for estimating the desirability of outcome ranking (DOOR) probability for benefit-risk evaluation in randomized trials and observational studies. The framework expresses the DOOR pr...
92. Agentic self-driving microscopy benchmarks support qualification but do not necessarily generalize to unseen tasks ​
Author: Nathan S Johnson, Ian Abshire
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cond-mat.mtrl-sci, cs.LG
arXiv:2608.05266v1 Announce Type: cross Abstract: Large language model agents are increasingly being developed to control a wide range of scientific characterization tools including microscopes and synchrotron beamlines. Research into agentic control of physical infrastructure is nascent and there a...
93. EdgeXpert: An Edge Device for Memory-Efficient LLM Inference with Mixture-of-Experts and Speculative Decoding ​
Author: Sangwoo Ha, Hyunwoo Seo, Yurim Jo, Youngjin Moon, Hoi-Jun Yoo
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AR, cs.CL, cs.LG
arXiv:2608.05303v1 Announce Type: cross Abstract: On-device deployment of Large Language Models (LLMs) has become essential for personalized edge applications. A primary bottleneck is external memory access (EMA) in feed-forward network (FFN) layers. Speculative decoding and mixture-of-experts (MoE)...
94. Failing Gracefully: Mitigating Impact of Inevitable Robot Failures ​
Author: Duc M. Nguyen, Saad A. Ghani, Andrew Marshall, Allison Andreyev, Gregory J. Stein, Xuesu Xiao
Published: 8/7/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.HC, cs.LG
arXiv:2608.05313v1 Announce Type: cross Abstract: Service robots operate in household environments shared with humans, pets, and everyday objects, where they are highly susceptible to failures such as software crashes, hardware degradation, or unpredictable interactions. While roboticists strive to ...
95. Computationally Efficient Collaborative Communication Via Regularity-Based Coarsening ​
Author: Mark Bedaywi, Scott Emmons, Nika Haghtalab, Stuart Russell
Published: 8/7/2026, 4:00:00 AM
Categories: cs.GT, cs.DS, cs.LG
arXiv:2608.05327v1 Announce Type: cross Abstract: Our results show that the existence of a short high-utility protocol already suffices for efficient communication. In particular, in a game with $n$ possible observations and $m$ actions: (1) For any achievable target utility $\alpha$, we give an alg...
96. Physics-Based Molecular Fingerprints from Spectral Graph Theory Provide Efficient Geometry-Aware Measures of Chemical Similarity ​
Author: Jacob W. Toney, Ayleen Y. Farnood, Samir Darouich, Heather J. Kulik
Published: 8/7/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG
arXiv:2608.05336v1 Announce Type: cross Abstract: Molecular representations are essential for the evaluation of molecular similarity and the development of structure-property relationships. Despite the known importance of 3D structure to determine chemical and physical properties, the most widely us...
97. Positive-Unlabeled Preference Optimization For Chest X-ray Report Generation ​
Author: Yuta Kobayashi, Pradyun Ramesh, Muhammad Ahmed Chaudhry, Vincent Jeanselme, Judy Wawira Gichoya, Sanmi Koyejo, Kathleen Capaccione, Shalmali Joshi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.05341v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) for radiology report generation are typically trained on retrospective clinical reports, which suffer from omission noise: clinically present findings are left unreported due to the omission of subtle findings. For examp...
98. Velocity- and Regime-Aware Detection of Intraday Options Market Manipulation, with Explainable Attribution ​
Author: Alex Chen, Maria Hybinette
Published: 8/7/2026, 4:00:00 AM
Categories: q-fin.TR, cs.LG, q-fin.ST
arXiv:2608.05373v1 Announce Type: cross Abstract: Intraday market manipulation is hard to detect because its footprint is brief, buried in millions of quotes, and statistically similar to ordinary volatility. Detectors reach high recall only by flagging so many other days that measured precision col...
99. DoctorAgents: an agentic framework to iteratively refine AutoML pipeline for small clinical temporal data ​
Author: Ruilin Wang, Bo-Hong Wang, Elizabeth Kourbatski, Jun Bai, Hegang Chen, Ziyang Song, Gilles Boire, Marie Hudson, Yue Li
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA
arXiv:2608.05375v1 Announce Type: cross Abstract: Clinical machine learning (ML) has the potential to support high-stakes medical decision-making, but reliable deployment is often constrained by scarce, heterogeneous, and temporal complexity. Developing effective ML pipelines for such data remains t...
100. Can Open-Weight LLMs Produce Kernel-Verified Coq Proofs? A Pilot Study ​
Author: Ahmed Ryan, Md Erfan, Akond Ashfaque Ur Rahman, Md Rayhanur Rahman
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LO, cs.LG
arXiv:2608.05420v1 Announce Type: cross Abstract: Large language models (LLMs) can generate text that resembles a mathematical proof, but resemblance does not establish correctness. A formal proof checker verifies whether each proof step follows established logical rules. Coq bases its rules on the ...
101. Invisible Shortcuts: Why Vision Encoders Know Your Camera ​
Author: Vladan Stojni'c, Ryan Ramos, Giorgos Kordopatis-Zilos, Noa Garcia, Giorgos Tolias
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.05424v1 Announce Type: cross Abstract: Deep vision models exploit shortcuts, relying on cues that correlate with supervision signals. Prior work has focused on visible biases, such as object-background or texture correlations. We identify a different source of shortcut learning: invisible...
102. Robust Context-Aware Detection of Malicious Instructions in Text ​
Author: Buzhao Liu, Xinhang Ma, Yevgeniy Vorobeychik
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.05430v1 Announce Type: cross Abstract: The remarkable instruction-following ability of modern LLMs has enabled their practical use as the minds of agents that can autonomously complete increasingly complex tasks. Therein, however, also lies their vulnerability to attacks which embed malic...
103. Discrete energy as an exact label-free training objective for finite-element surrogates ​
Author: Ruifeng Cao (The University of Manchester), Xidan Song (Wuhan University)
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CE, cs.LG, cs.NA, math.NA
arXiv:2608.05437v1 Announce Type: cross Abstract: Supervised training of finite-element (FE) surrogate models requires reference solutions, and each reference solution is obtained by solving the system that the surrogate is intended to replace. The assembled discrete potential energy provides a trai...
104. SCP-NL2TL: Selective Conformal Prediction with Semantic Verification for Natural Language to Temporal Logic Specifications ​
Author: Yixuan Wang, Licheng Luo, Yu Fu, Kaidi Xu, Yue Dong, Mingyu Cai
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.05439v1 Announce Type: cross Abstract: Translating natural language instructions into machine-interpretable formal specifications enables robots and autonomous systems to plan, reason, and formally verify their behavior. However, existing translation models typically generate a specificat...
105. Effective pruning of task-trained recurrent neural networks using noisy fluctuations and connection rescaling ​
Author: Sanjith Senthil, Rishidev Chaudhuri
Published: 8/7/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG, cs.NE
arXiv:2608.05464v1 Announce Type: cross Abstract: The pruning of network connections is key to brain function but, despite its importance, there exist few biologically-plausible pruning rules with demonstrated good performance. In this work we evaluate noise-prune, a recently introduced unsupervised...
106. Recursive Synthesis for Long-Horizon Terminal Tasks ​
Author: Zhongzhi Li, Yucheng Shi, Zongxia Li, Ruhan Wang, Anhao Li, Zixun Huang, Junyao Yang, Lei Ke, Ninghao Liu, Haitao Mi, Leowei Liang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.05466v1 Announce Type: cross Abstract: High-quality long-horizon training data for terminal agents is expensive to produce, often costing hundreds to thousands of dollars per task, because each task must keep the instruction, environment, reference solution, and verifier mutually consiste...
107. A Foundational EDM2-Based Generative Model for High-Resolution Synthetic Fetal Ultrasound Imaging from Open Datasets ​
Author: Harvey Mannering, Yilin Zhang, Ziao Liu, Zhiwu Huang, Jacqueline Matthew, Miguel Xochicale
Published: 8/7/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, physics.med-ph
arXiv:2608.05471v1 Announce Type: cross Abstract: Prenatal ultrasound imaging is key for assessing fetal health, but AI progress is limited by scarce, privacy-restricted, and hard-to-annotate datasets. We propose a high-resolution fetal ultrasound synthesis framework based on the EDM2 diffusion arch...
108. GenGA: Editable and Data-Grounded Graphical Abstract Generation for Academic Papers ​
Author: Takuro Kawada, Shunsuke Kitada, Hitoshi Iyatomi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.GR, cs.CL, cs.CV, cs.HC, cs.LG, cs.MM
arXiv:2608.05478v1 Announce Type: cross Abstract: Graphical Abstracts (GAs) visually summarize the key findings of academic papers, playing a crucial role in facilitating the understanding of research content. Recently, advancements in vision-language models and image generation models have enabled ...
109. Innovation-Residual Auditing of Autonomous Analysis Agents: Localization, Detection Limits, Error Control, and Identifiability ​
Author: Ahmed Hassoon, Mark Dredze
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, stat.ML
arXiv:2608.05490v1 Announce Type: cross Abstract: Autonomous agents now carry out entire data analyses, selecting cohorts, joining tables, and fitting models with little step-by-step supervision. When such an analysis turns out to be wrong, someone must determine which operation caused it. A recent ...
110. APQF: Agentic Profiling-Guided Structured Pruning and Mixed-Precision Quantization with Adaptive Fine-Tuning ​
Author: Sadegh Jafari, Mohiuddin Bilwal, Fan Zhou, Brian Gelder, Ali Jannesari
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.05499v1 Announce Type: cross Abstract: Modern deep neural networks achieve strong performance, but their scale makes them costly and slow, especially on resource-constrained edge devices. Pruning and quantization address this, but rely on manual, expert choices and on algorithms that are ...
111. An Inertial Block Proximal Linearized Method with Adaptive Momentum for Nonconvex and Nonsmooth Optimization ​
Author: Weifeng Yang
Published: 8/7/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2608.05502v1 Announce Type: cross Abstract: In this paper, we consider a class of multiblock nonconvex nonsmooth optimization problems, which covers many applications such as the analysis of pre-earthquake anomalies and machine learning. To solve this class of problems, we propose the inertial...
112. EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents ​
Author: Jie Wu, Ming Gong, Feixiang Cheng, Qinqin Zhao
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.05519v1 Announce Type: cross Abstract: Agent benchmarks usually measure task completion and treat resource use as an auxiliary statistic. In deployment, however, the choice among a local lookup, broad search, composite research tool, stronger model, or human escalation is part of the task...
113. Equation-Free Period-Aware Forecast-Error Contraction for Estimating Negative Largest Lyapunov Exponents from Short Trajectory Ensembles ​
Author: Andrei Velichko, N'Gbo N'Gbo, Viet-Thanh Pham
Published: 8/7/2026, 4:00:00 AM
Categories: nlin.CD, cs.LG, physics.data-an
arXiv:2608.05522v1 Announce Type: cross Abstract: Estimating positive largest Lyapunov exponents from data is comparatively natural because neighboring trajectories separate, whereas stable dynamics require resolving contraction before measurement noise or finite precision erases the signal. We intr...
114. Behavioral Residualization for Unsupervised Intrusion Detection in Automotive CAN Networks ​
Author: Chandan Hegde, Mukundh R Reddy
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.05548v1 Announce Type: cross Abstract: Modern vehicles rely on the Controller Area Network (CAN) bus, whose design prioritizes low cost and real-time performance but provides no message authentication or encryption. An attacker with physical or remote access can therefore inject arbitrary...
115. The Judgment-Consequence Gap: LLM Moral Reasoning in Healthcare Decisions ​
Author: Hadi Hosseini, Samarth Khanna, Leona Pierce
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG
arXiv:2608.05583v1 Announce Type: cross Abstract: As large language models (LLMs) enter high-stakes domains such as healthcare, understanding their moral reasoning becomes essential. Decisions about scarce medical resources often hinge on judgments of responsibility, particularly when patients' own ...
116. How Much Reconstruction Does Quantum Machine Learning Need? Late Fusion of Independently Trained Quantum Subcircuits ​
Author: Prabhjot Singh, Adel N. Toosi, Rajkumar Buyya
Published: 8/7/2026, 4:00:00 AM
Categories: quant-ph, cs.DC, cs.LG
arXiv:2608.05595v1 Announce Type: cross Abstract: Circuit cutting lets a large quantum neural network (QNN) run as independent subcircuits on small devices, but rebuilding its outputs by reconstruction carries a classical sampling overhead exponential in the number of cuts - the dominant runtime cos...
117. Enhancing Anomaly Resilience in Research Networks: A Large-Scale Forecasting Benchmark for Dynamic Security Baselining ​
Author: Mohammad Arafath Uddin Shariff, Byrav Ramamurthy
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.NI
arXiv:2608.05605v1 Announce Type: cross Abstract: Research and Education Networks (RENs) serve as critical infrastructure for scientific discovery, yet they face a unique security paradox: their normal traffic patterns which are characterized by massive, bursty "elephant flows" are statistically ind...
118. LC-Implicit-QAOA: Active-Workspace-Capped Exact Objective-and-Gradient Evaluation for Training over Bounded QUBO Light Cones ​
Author: Chih-Chung Hsu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.ET, cs.LG
arXiv:2608.05610v1 Announce Type: cross Abstract: QAOA training repeatedly queries an objective and all shared gradients, making exact evaluation a feasibility bottleneck even when QUBO terms have bounded causal cones. Building on established causal-cone restriction and adjoint differentiation, LC-I...
119. FOCUS: Decoupling Expert Personas in LLMs to Enhance Domain Expert Capabilities ​
Author: Guanyu Wang, Zidi Zhang, Xu Chu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.05611v1 Announce Type: cross Abstract: Large Language Models (LLMs) can exhibit diverse personas, and activating expert personas has been shown to improve domain expertise and task accuracy. However, existing persona control methods often suffer from cross-domain coupling, which may lead ...
120. RASP-QAOA: Resource-Aware Per-Instance Selection for Exact QAOA Simulation ​
Author: Chih-Chung Hsu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.ET, cs.LG
arXiv:2608.05646v1 Announce Type: cross Abstract: Exact QAOA simulation spans several computational representations whose useful regions differ sharply across graph structure, circuit depth, precision, and available memory. Choosing only a backend name hides these differences: an executable choice a...
121. A Unified Framework for Trajectory Prediction with Explicit Planning and Reaction Decomposition ​
Author: Jiaheng Chen, Jiaxing Li, Tinghe Zhang, Chaopeng Guo
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG
arXiv:2608.05673v1 Announce Type: cross Abstract: Trajectory prediction has shifted toward structured formulations with explicit social modeling. However, existing methods inadequately distinguish the functional roles of social influence in trajectory planning. Observing that agents typically form m...
122. Nonvisual Classification of Ground-Condition by Artificial Proprioception in an Amoeba-Inspired Autonomous Walking Robot ​
Author: Hyoto Yamaguchi, Zenji Yatabe, Seiya Kasai
Published: 8/7/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, cs.SY, eess.SY
arXiv:2608.05684v1 Announce Type: cross Abstract: Nonvisual classification of ground condition based on a multimodal sensing approach was investigated for an amoeba-inspired autonomous walking robot. To classify ground condition without image sensing and processing, we implemented artificial proprio...
123. Provably Efficient Self-Calibrating Quantum Fault Tolerance ​
Author: Weiyuan Gong, Hong-Ye Hu
Published: 8/7/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.05686v1 Announce Type: cross Abstract: Quantum error correction protects logical information only when every physical operation remains below the fault-tolerance threshold, a condition that must be maintained continuously rather than only at the initial calibration. In practice, however, ...
124. A Low-Power Wearable Respiratory Sensor for Non-Invasive Stress Monitoring ​
Author: Mohammad Hosseini, Hamed Khatounabadi, Mohammad Fakharzadeh
Published: 8/7/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG, eess.SP
arXiv:2608.05697v1 Announce Type: cross Abstract: Respiration provides a continuously available window into physiological state and behavior. However, monitoring it outside controlled settings remains challenging because a wearable system must capture small body deformations while remaining comforta...
125. Sparse Mutual Information Graph Averaging for Improving Random Indexing Embeddings ​
Author: Sriram Loganathan, Gokul Anand, Aung Bo Bo, Yourui Shao, William B. Andreopoulos
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.05724v1 Announce Type: cross Abstract: Sparse word embedding pipelines can avoid dense co-occurrence matrix materialization, dense factorization, and gradient training while still relying on sparse global corpus statistics. This paper studies Random Indexing (RI) vectors refined by weight...
126. LILAC: An Idempotent Neural Speech Codec ​
Author: June Young Yi, Dongwook Lee, Jiheum Yeom, Sungroh Yoon
Published: 8/7/2026, 4:00:00 AM
Categories: cs.SD, cs.LG
arXiv:2608.05727v1 Announce Type: cross Abstract: Neural Audio Codecs are widely adopted in speech generation and editing. However, existing neural audio codecs are not idempotent: across the paper's twelve baseline systems, every configuration tested rewrites, on average, at least 15% of its tokens...
127. Engram-E2VID: Reference-Based Event-to-Video Reconstruction via Generative Activation of Appearance Engrams ​
Author: Feiyu Ji, Xiang Li, Hao Ma, Tianxiang Huang, Qingxin Lu, Mengqi Ji, Lei Han, Xiaokang Yang, Xiaoyun Yuan
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, physics.optics
arXiv:2608.05728v1 Announce Type: cross Abstract: Reference-based event-to-video reconstruction aims to recover target RGB frames from a reference frame and the event stream captured over the reference-to-target interval. Although events provide fine-grained temporal cues, they encode sparse and asy...
128. ABC: Numerical Data Collection under Local Differential Privacy without Prior Knowledge ​
Author: Incheol Baek, Hyungbin Kim, Yon Dohn Chung
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2608.05737v1 Announce Type: cross Abstract: Local Differential Privacy (LDP) provides strong privacy guarantees for collecting numerical data. A fundamental challenge, however, is that existing LDP mechanisms require a predefined data domain, which is often unknown in practice. This lack of pr...
129. SR-JEPA: Learning Predictive Latent State in 3D Scenes ​
Author: Zihan Zhou, Qifu Wen, Xi Zeng
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.05774v1 Announce Type: cross Abstract: Joint-embedding predictive architectures learn by predicting latent representations of missing observations, yet many masked JEPAs are evaluated primarily through the encoders they produce. We ask what a trained predictive pathway itself infers when ...
130. VSMP-IMU: Video-Grounded Semantic Motion Programs for Sensor-Aware Synthetic IMU Generation ​
Author: Lala Shakti Swarup Ray, Vitor Fortes Rey, Mengxi Liu, Paul Lukowicz, Bo Zhou
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.05782v1 Announce Type: cross Abstract: Wearable human activity recognition (HAR) is often limited by the scarcity of labeled sensor data, especially in low-resource, class-imbalanced, and subject-generalization settings. Synthetic IMU generation can reduce this dependency and enhance HAR ...
131. KVAE: Family of Tokenizers for Multimodal Generative Models ​
Author: Andrey Shutkin, Denis Parkhomenko, Ivan Kirillov, Kirill Chernyshev, Kirill Malakhov, Ilia Vasiliev, Ilia Trushkin, Valeriya Kobenko, David Chikovani, Alexander Ivanov, Azat Saginbaev, Egor Silvestrov, Ivan Mikheev, Konstantin Zakharov
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.SD
arXiv:2608.05798v1 Announce Type: cross Abstract: Latent diffusion modeling (LDM), a prominent paradigm, utilizes tokenizers to map input signal to compressed representation. This dependency positions tokenizer as an integral part of generation process itself, since it affects learning speed, qualit...
132. On-Policy Delta Distillation for Multilingual Math Reasoning ​
Author: Byeongho Heo, Jaehui Hwang, Sangdoo Yun, Dongyoon Han
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.05802v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) is emerging as a promising alternative to reinforcement learning for LLM post-training, yet its effectiveness in multilingual settings remains underexplored. We study OPD and its advanced variant, On-Policy Delta Distilla...
133. A neural operator view on U-Nets for inverse imaging problems ​
Author: Alexander Auras, Martin Burger, Samira Kabri, Michael Moeller, Michael Schopf-Kuester
Published: 8/7/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA
arXiv:2608.05839v1 Announce Type: cross Abstract: Deep neural networks have shown great empirical success in the solution of a wide variety of ill-posed inverse problems in imaging. Yet, very few works have studied their behavior in the limit that turns the discretized ill-conditioned problems into ...
134. Deep Generalised Mixed Models: a Novel Neural Network Structure for Analysing Hierarchical Data ​
Author: Nina van Gerwen, Dimitris Rizopoulos, Manon Hillegers, Loes Keijsers, Sten Willemsen
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.05930v1 Announce Type: cross Abstract: The experience sampling method (ESM) is a longitudinal research design where participants report their thoughts, emotional states and behaviours multiple times a day. Our work is motivated by such data collected by the GrowIt! app, which was released...
135. MirrorNet: Can Medical Image Anonymization Really Protect Patient Identity? ​
Author: Attila Simk'o
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.05938v1 Announce Type: cross Abstract: Medical images are routinely de-identified---names, dates, and other metadata removed---and then shared for research, teaching, and public benchmarks under the assumption that this renders them anonymous. Such de-identification protects the metadata ...
136. Operating Multi-Node Full Fine-Tuning on NVIDIA B300: A Field Report on Telemetry-Based Triage, Negative Results, and Operational Hardening ​
Author: Seon Ho Kim, Ui Jeong Jeon, Su Hyeon Kim, Min Tae Hwang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2608.05944v1 Announce Type: cross Abstract: We report operational experience full-fine-tuning a 32.76B-parameter dense model (Qwen3-32B) on 16 x NVIDIA B300 (two nodes, FSDP / ZeRO-3) -- among the first published field accounts on this accelerator. We claim no new algorithm. The individual mec...
137. VLMs for Videogame Data Annotation ​
Author: Katrin Schmid, Iuri Frosio
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG
arXiv:2608.05949v1 Announce Type: cross Abstract: Vision Language Models (VLMs) and Artificial Intelligence (AI) agents have revolutionized how engineers approach complex problems in real-world applications. Their adoption in video games is on the other hand limited by the extreme variability of the...
138. Training a Conditioned Video Game Agent on a VLM Annotated Dataset ​
Author: Katrin Schmid, Iuri Frosio
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG
arXiv:2608.05954v1 Announce Type: cross Abstract: Reinforcement Learning (RL) is a powerful but far from easy-to-use technique for policy learning. In the specific case of video games, access to the game engine is required to get rewards for training (e.g. to collect rewards from the environment). F...
139. Temporal Bridges for Spatial Resolution: Enhancing Climate Data Super-Resolution with Bidirectional Alignment ​
Author: Yichen Zhang, Yixiong Xiao, Congxi Xiao, Jingbo Zhou
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.05981v1 Announce Type: cross Abstract: High-resolution climate data is crucial for meteorological predictions and for informing decision support across diverse domains. However, the acquisition of such high-resolution climate information is often prohibitively costly, necessitating the de...
140. AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning ​
Author: Zi-Han Wang, Zhengxi Lu, Zhiyuan Yao, Jinyang Wu, Jie Wu, Zhengzhou Cai, Yueqing Sun, Ziang Ye, Linji Hao, Qi Gu, Xunliang Cai, Yongliang Shen, Yujiu Yang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.05987v1 Announce Type: cross Abstract: Reinforcement learning (RL) with verifiable rewards constructs trajectory-level advantage estimates, yet it often fails to credit the few pivotal decisions that determine outcomes in long-horizon, multi-turn agentic tasks. Recent work introduces priv...
141. From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models ​
Author: Jiale Han, Xiang Li, Jing Qian, Wenyuan Gu, Pin Gao, Ye Luo, Hongyuan Zha, Dacheng Tao, Benyou Wang, Lin William Cong
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.06020v1 Announce Type: cross Abstract: Economic World Models (EWMs) are generative economic models that simulate how economies evolve from within by modeling heterogeneous agents, their beliefs and actions, and the market and institutional mechanisms through which their interactions produ...
142. Integrating Implicit and Explicit Relational Biases through Graph-Based Multiple Instance Learning: A Case Study in Skin Lesion Diagnosis ​
Author: Rafa{\l} Buler (Gda'nsk University of Technology), Jakub Buler (Gda'nsk University of Technology), Maciej Bobowicz (Medical University of Gda'nsk), Micha{\l} Grochowski (Gda'nsk University of Technology)
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG, eess.IV
arXiv:2608.06037v1 Announce Type: cross Abstract: Relational inductive biases are essential for capturing structural dependencies among data. This study investigates a dual-level relational framework for image classification, bridging the gap between implicit representation learning and explicit str...
143. ML-for-ML ​
Author: Yutong Zhao, Noga H. Rotman, Gianni Antichi, Ran Ben Basat
Published: 8/7/2026, 4:00:00 AM
Categories: cs.NI, cs.DC, cs.LG
arXiv:2608.06046v1 Announce Type: cross Abstract: AI training workloads are growing rapidly, making their time, energy, and infrastructure costs increasingly important. In shared cloud clusters, training and fine-tuning jobs compete with co-running workloads for network resources, while network mech...
144. From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems ​
Author: Manideep Dhar, Ritwik Singh, Sharat Chandra Kumar Manikonda
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA
arXiv:2608.06112v1 Announce Type: cross Abstract: Hospitals are rapidly adopting artificial intelligence for triage, imaging, scheduling etc., yet most deployments remain isolated point solutions locked inside departmental silos, resulting in duplicated effort, hidden risks, and unrealized enterpris...
145. Hardware Keystores for AI Agent Signing Workflows: A Zero-Trust MCP Enforcement Architecture ​
Author: Leo Sambrook, Sampo Sovio
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2608.06130v1 Announce Type: cross Abstract: AI agents performing cryptographic operations (signing Git commits, authenticating API calls, issuing certificates) currently store private keys in software-accessible locations: plaintext files, environment variables, or container memory. Any proces...
146. Verifiable Regularity Criterion for Conditional Expectation Operators and Conditional Mean Embeddings with Applications to Nonparametric Regression, Bayesian Inverse Problems, and Koopman Operators ​
Author: Maximiliano Hertel, Ilja Klebanov, Manuel Schaller, Karl Worthmann
Published: 8/7/2026, 4:00:00 AM
Categories: math.DS, cs.LG, cs.NA, math.NA, math.ST, stat.ML, stat.TH
arXiv:2608.06155v1 Announce Type: cross Abstract: Conditional expectation operators (CEOs) and their associated conditional mean embeddings (CMEs) play a central role across applied mathematics and machine learning, appearing in nonparametric regression, Bayesian inverse problems, and Koopman operat...
147. On Same-Sample and Independent-Sample Stochastic Extragradient for Monotone Variational Inequalities ​
Author: TaeHo Yoon, Nicolas Loizou
Published: 8/7/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2608.06182v1 Announce Type: cross Abstract: We study stochastic extragradient (SEG) methods for solving monotone variational inequality problems (VIPs) over a feasible set. Although extragradient is a foundational algorithm for VIPs and its deterministic convergence theory is well developed, i...
148. Handling Missing Data in Probabilistic Regression Trees ​
Author: Taiane Schaedler Prass, Alisson Silva Neimaier, Guilherme Pumi
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.06195v1 Announce Type: cross Abstract: Probabilistic Regression Trees (PRTrees) are a smooth and consistent alternative to classical regression trees, producing continuous predictions through probabilistic split assignments. This paper extends the PRTree framework to accommodate missing p...
149. Beyond Marginal Validity: Finite-Sample Guarantees for Localized Conformal Prediction ​
Author: Anton Conrad, Rustam Isaev, Denis Belomestny, Eric Moulines, Sergey Samsonov
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.06206v1 Announce Type: cross Abstract: Conformal prediction endows arbitrary black-box predictors with finite-sample, distribution-free marginal coverage, yet marginal validity can hide severe covariate-specific miscalibration, while exact distribution-free conditional coverage is finite-...
150. Muon on the Stiefel Manifold Admits an Exact Closed-Form Update ​
Author: Mikhail Solonko, Molozhavenko Alexander, Maxim Rakhuba
Published: 8/7/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.NA, math.NA
arXiv:2608.06218v1 Announce Type: cross Abstract: We study Muon, a recently proposed matrix-aware optimization method, in the context of the Stiefel manifold. This manifold consists of matrices with orthonormal columns and is ubiquitous in machine learning and scientific computing. Existing extensio...
151. Robot Learning from Human Demonstrations: Handwritten Alphabet Trajectories and Human-Likeness Evaluation ​
Author: Alperen Kenan, Paul Bremner, Manuel Giuliani
Published: 8/7/2026, 4:00:00 AM
Categories: cs.RO, cs.HC, cs.LG
arXiv:2608.06221v1 Announce Type: cross Abstract: Learning from demonstration (LfD) provides a developmental framework through which robots can develop motor skills by observing and imitating human dynamics, reducing reliance on explicit programming to teach a skill to a robot. The resulting human-l...
152. TS-RAG: Retrieval Augmented Generation for Time Series Forecasting ​
Author: Yixiong Xiao, Congxi Xiao, Jingbo Zhou
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.06223v1 Announce Type: cross Abstract: While deep learning models, particularly transformer-based architectures, have shown impressive performance in time series forecasting, the application of retrieval-augmented generation (RAG) in this domain remains limited. Since RAG has proven effec...
153. Minimax Optimal Early-Stopped Gradient Descent for Gaussian Mixture Classification ​
Author: Alex Buna, Shirley Xiaoqi Liu, Patrick Rebeschini
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.06250v1 Announce Type: cross Abstract: In overparameterised classification, training data can be linearly separable even when the underlying distribution is not. In this setting, gradient descent (GD) on the logistic loss diverges in norm while converging in direction to a max-margin inte...
154. OTLesMix: Wasserstein Barycenter and Optimal Transport Map for Synthetic Lesion Generation with Diverse Shapes and Locations ​
Author: Robin Trombetta, Carole Lartizien
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV
arXiv:2608.06264v1 Announce Type: cross Abstract: The development of deep learning over the past decade has revolutionized medical imaging segmentation, allowing the extraction of precise descriptors from large volumes to characterize pathologies. Data augmentation is a technique widely regarded as ...
155. Improving the Realism of Synthetic Clinical Benchmarks Under Utility Constraints ​
Author: Omid Bazgir, Md Nasir, Jacob Hoffman, Yang Yang, Manu Agrawal, Anusua Trivedi, Vinay Rao Dandin, Chris Gibbons, Christine Swisher
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.DB, cs.LG
arXiv:2608.06265v1 Announce Type: cross Abstract: Synthetic clinical benchmarks for enterprise AI agents can pass existing utility checks and still remain structurally unrealistic, especially in privacy-sensitive healthcare settings where operational data are hard to access. We study how to improve ...
156. Stochastic Dynamics on Persistence Diagram Space via Reinforcement Learning ​
Author: Farzana Nasrin
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.AT
arXiv:2608.06276v1 Announce Type: cross Abstract: Persistence diagrams (PDs) provide stable and interpretable summaries of multiscale topological structure. While substantial progress has been made in the statistical analysis of PDs, existing literature often treats diagrams as static objects and pr...
157. HarnessOpt-Bench: Evaluating LLMs at Harness Optimization ​
Author: Varun Ursekar, Apaar Shanker, Yash Maurya, Shehab Yasser, Vijay S. Kalmath, Veronica Chatrath, Yuan Xue
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.06301v1 Announce Type: cross Abstract: As LLMs are increasingly deployed within agentic systems, their capabilities depend not only on the model weights but also on the harness: the prompts, tools, control flow, memory, and orchestration code surrounding them. This makes automated harness...
158. Optimal Rates for Learning with Monotone Adversaries ​
Author: Anay Mehrotra
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.DS, cs.LG, math.ST, stat.TH
arXiv:2608.06337v1 Announce Type: cross Abstract: A monotone adversary observes an i.i.d. labeled sample and appends a finite number of further examples of its choice, every one of them labeled correctly by the target hypothesis. The learner sees a uniform shuffle of the combined sample and is score...
159. Scalable estimation of VARMA models ​
Author: Daniel Paulin, Victor Elvira
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2608.06340v1 Announce Type: cross Abstract: Vector autoregressive moving-average (VARMA) models have long been considered impractical beyond moderate dimensions: the likelihood is non-convex, the parametrization is identified only up to equivalence, and every evaluation costs a pass over the e...
160. AV-AIVAT: 74x Cheaper Agent Evaluation with Certified Anytime-Valid Stopping in Imperfect-Information Games ​
Author: Boning Li, Yu Chen, Longbo Huang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.GT, cs.AI, cs.CL, cs.LG, cs.MA
arXiv:2608.06362v1 Announce Type: cross Abstract: Deciding which of two agents is stronger means playing games until skill outweighs luck, and every game costs money, model inference, or expert time. Since the number of games needed is unknown, fixed-budget evaluations either keep paying after the r...
161. Tracing the Heart: An Evidence-Linked Pipeline for Heart-Failure Feature Engineering ​
Author: Soorya Ram Shimgekar, Michelle Hu, Dorisa Shehi, Daniel Kang, Roy Ka-Wei Lee, Koustuv Saha, Christian Poellabauer, Christopher Lee, Sajeev Singh, Piyum Zonooz, Navin Kumar, Zeeshan Ahmed, Priyadarshini Kachroo
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.06366v1 Announce Type: cross Abstract: Electronic health record (EHR) feature engineering is a major bottleneck in clinical research and AI, accounting for 39-45% of data scientists' workload. This is especially pronounced in heart failure, which affects an estimated 6.7 million U.S. adul...
162. Learning When to Trust via Selective Context Preference Optimization ​
Author: Xian Sun, Wei Chow, Yingshuo Wang, Junhao Liu, Wei Gao, Qing Wu, Lingdong Kong
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.06377v1 Announce Type: cross Abstract: Language models increasingly condition their answers on external signals, and a single misleading one can turn a correct answer wrong. The obvious remedy, training models to resist such signals, hides a failure mode: a model that ignores all context ...
163. Optimal or Greedy Decision Trees? Revisiting their Objectives, Tuning, and Performance ​
Author: Jacobus G. M. van der Linden, Dani"el Vos, Mathijs M. de Weerdt, Sicco Verwer, Emir Demirovi'c
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2409.12788v3 Announce Type: replace Abstract: Recently there has been a surge of interest in optimal decision tree (ODT) methods that globally optimize accuracy directly, in contrast to traditional approaches that locally optimize an impurity or information metric. However, the literature show...
164. ASAT: Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection ​
Author: Daisuke Yamada, Harit Vishwakarma, Ramya Korlakai Vinayak
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2505.02299v2 Announce Type: replace Abstract: Machine Learning (ML) models are trained on in-distribution (ID) data but often encounter out-of-distribution (OOD) inputs during deployment---posing serious risks in safety-critical domains. Recent works have focused on designing scoring functions...
165. Scientific Machine Learning of Chaotic Systems Learns Reduced-Order Equations for Neural Populations ​
Author: Anthony G. Chesebro, David Hofmann, Vaibhav Dixit, Earl K. Miller, Richard H. Granger, Alan Edelman, Christopher V. Rackauckas, Lilianne R. Mujica-Parodi, Helmut H. Strey
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, nlin.CD, q-bio.NC
arXiv:2507.03631v4 Announce Type: replace Abstract: Extracting interpretable mathematical models from complex dynamical systems is difficult, especially for chaotic dynamics observed with noisy experimental data. We present PEM-UDE, a method that combines prediction-error methodology with universal ...
166. CRINN: Contrastive Reinforcement Learning for Approximate Nearest Neighbor Search ​
Author: Xiaoya Li, Albert Wang, Guoyin Wang, Chris Shum, Jiwei Li
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.DB
arXiv:2508.02091v4 Announce Type: replace Abstract: Approximate nearest-neighbor search (ANNS) algorithms have become increasingly critical for recent AI applications, particularly in retrieval-augmented generation (RAG) and agent-based LLM applications. In this paper, we present CRINN, a new paradi...
167. Uncertainty-aware Predict-Then-Optimize Framework for Equitable Post-Disaster Power Restoration ​
Author: Lin Jiang, Dahai Yu, Rongchao Xu, Tian Tang, Guang Wang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SI
arXiv:2508.04780v3 Announce Type: replace Abstract: The increasing frequency of extreme weather events, such as hurricanes, highlights the urgent need for efficient and equitable power system restoration. Many electricity providers make restoration decisions primarily based on the volume of power re...
168. HCRide: Harmonizing Passenger Fairness and Driver Preference for Human-Centered Ride-Hailing ​
Author: Lin Jiang, Yu Yang, Guang Wang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2508.04811v3 Announce Type: replace Abstract: Order dispatch systems play a vital role in ride-hailing services, which directly influence operator revenue, driver profit, and passenger experience. Most existing work focuses on improving system efficiency in terms of operator revenue, which may...
169. Autonomous Learning From Success and Failure: Goal-Conditioned Supervised Learning with Negative Feedback ​
Author: Zeqiang Zhang, Fabian Wurzberger, Gerrit Schmid, Sebastian Gottwald, Daniel A. Braun
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2509.03206v2 Announce Type: replace Abstract: Learning from reward functions and imitation learning of demonstrations are the two principal approaches for training autonomous systems that interact with an environment through action and observation. Both, however, require human specification fo...
170. Scaling Laws and Spectra of Shallow Neural Networks in the Feature Learning Regime ​
Author: Leonardo Defilippis, Yizhou Xu, Julius Girardin, Emanuele Troiani, Vittorio Erba, Lenka Zdeborov'a, Bruno Loureiro, Florent Krzakala
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, cs.AI, stat.ML
arXiv:2509.24882v3 Announce Type: replace Abstract: Neural scaling laws underlie many of the recent advances in deep learning, yet their theoretical understanding remains largely confined to linear models. In this work, we present a systematic analysis of scaling laws for quadratic and diagonal neur...
171. Invariant Representation Learning for Source-Free Time Series Forecasting with LLM-Centric Proxy Denoising ​
Author: Kangjia Yan, Chenxi Liu, Hao Miao, Xinle Wu, Yan Zhao, Chenjuan Guo, Bin Yang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.05589v3 Announce Type: replace Abstract: Effective time series forecasting enables various real-world applications, benefiting from the proliferation of mobile devices. However, the volume of time series data may vary significantly across domains due to high data acquisition costs and dat...
172. Worst-Case Distance-Aware Error Bounds for Neural Networks ​
Author: Masoud Ataei, Vikas Dhiman, Mohammad Javad Khojasteh
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, stat.ML
arXiv:2510.22021v3 Announce Type: replace Abstract: Safety-critical applications of machine learning require uncertainty estimates that support reliable worst-case analysis. Neural networks (NNs) provide expressive function approximation, while Gaussian processes (GPs) offer principled probabilistic...
173. On the Anisotropy of Score-Based Generative Models ​
Author: Andreas Floros, Seyed-Mohsen Moosavi-Dezfooli, Pier Luigi Dragotti
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2510.22899v2 Announce Type: replace Abstract: We investigate the role of network architecture in shaping the inductive biases of modern score-based generative models. To this end, we introduce the Score Anisotropy Directions (SADs), architecture-dependent directions that reveal how different n...
174. Trajectory-guided discharge stratification for heart failure using short-context electronic health record sequence modeling ​
Author: Falk Dippel, Yinan Yu, Annika Rosengren, Martin Lindgren, Christina E. Lundberg, Erik Aerts, Martin Adiels, Helen Sj"oland
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2511.16839v4 Announce Type: replace Abstract: Purpose: Heart failure (HF) discharge planning depends on identifying patients at risk of deterioration or death, yet accurate prediction from routinely collected electronic health records (EHRs) remains challenging. Methods: We develop trajectory-...
175. Transformers with RL or SFT Provably Learn Sparse Boolean Functions, But Differently ​
Author: Bochen Lyu, Yiyang Jia, Xiaohao Cai, Zhanxing Zhu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2511.17852v3 Announce Type: replace Abstract: Transformers can acquire Chain-of-Thought (CoT) capabilities to solve reasoning tasks via fine-tuning. Reinforcement learning (RL) and supervised fine-tuning (SFT) are two primary approaches to this end. In this work, we examine RL with verifiable ...
176. CUDA-L2: Surpassing cuBLAS Performance for Matrix Multiplication through Reinforcement Learning ​
Author: Songqiao Su, Xiaoya Li, Albert Wang, Guoyin Wang, Jiwei Li, Chris Shum
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2512.02551v4 Announce Type: replace Abstract: In this paper, we propose CUDA-L2, a system that combines large language models (LLMs) and reinforcement learning (RL) to automatically optimize Half-precision General Matrix Multiply (HGEMM) CUDA kernels. Using CUDA execution speed as the RL rewar...
177. All-Quadrant Bounded Clipping GRPO: Closing the Unbounded Blind Spot for Stable and Generalizable Training ​
Author: Chi Liu, Xin Chen
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2601.03895v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has emerged as a popular algorithm for reinforcement learning with large language models (LLMs). However, GRPO inherits PPO's token-level clipping while replacing token-level advantages with a single sequen...
178. d3LLM: Ultra-Fast Diffusion LLM using Pseudo-Trajectory Distillation ​
Author: Yu-Yang Qian, Junda Su, Lanxiang Hu, Peiyuan Zhang, Zhijie Deng, Peng Zhao, Hao Zhang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2601.07568v3 Announce Type: replace Abstract: Diffusion large language models (dLLMs) offer capabilities beyond those of autoregressive (AR) LLMs, such as parallel decoding and random-order generation. However, realizing these benefits in practice is non-trivial, as dLLMs inherently face an ac...
179. FI-TW: An Open Train-Weather Dataset for Railway Delay Analysis in Finland ​
Author: Vinicius Pozzobon Borin, Jean Michel de Souza Sant'Ana, Usama Raheel, Nurul Huda Mahmood
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DB
arXiv:2601.16592v2 Announce Type: replace Abstract: Train delays result from complex interactions between operational, technical, and environmental factors. While weather impacts railway reliability, particularly in Nordic regions, existing datasets rarely integrate meteorological information with o...
180. On the Limits of Layer Pruning for Generative Reasoning in Large Language Models ​
Author: Safal Shrestha, Anubhav Shrestha, Aadim Nepal, Minwu Kim, Keith Ross
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.01997v4 Announce Type: replace Abstract: Recent work has shown that layer pruning can effectively compress large language models (LLMs) while retaining strong performance on classification benchmarks, often with little or no finetuning. In contrast, generative reasoning tasks, such as GSM...
181. Trust-Based Incentive Mechanisms in Semi-Decentralized Federated Learning Systems ​
Author: Ajay Kumar Shrestha
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET
arXiv:2602.08290v2 Announce Type: replace Abstract: In federated learning (FL), decentralized model training allows multi-ple participants to collaboratively improve a shared machine learning model without exchanging raw data. However, ensuring the integrity and reliability of the system is challeng...
182. Continuous-Time Piecewise-Linear Recurrent Neural Networks ​
Author: Alena Br"andle, Lukas Eisenmann, Florian G"otz, Daniel Durstewitz
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.15649v2 Announce Type: replace Abstract: In dynamical systems reconstruction (DSR) we aim to recover the dynamical system (DS) underlying observed time series. Specifically, we aim to learn a generative surrogate model which approximates the underlying, data-generating DS, and recreates i...
183. MARS: Margin and Semantic-Aware Data Augmentation for Reward Modeling ​
Author: Payel Bhattacharjee, Osvaldo Simeone, Ravi Tandon
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT
arXiv:2602.17658v4 Announce Type: replace Abstract: Reward modeling is central to RLHF, RLAIF, and PPO-based alignment, but its reliability is often limited by scarce and heterogeneous human preference data. In this paper, we introduce MARS (Margin and Semantic-Aware Data Augmentation for Reward Mod...
184. Dynamics of Learning under User Choice: Overspecialization and Peer-Model Probing ​
Author: Adhyyan Narang, Sarah Dean, Lillian J Ratliff, Maryam Fazel
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.MA
arXiv:2602.23565v3 Announce Type: replace Abstract: In many economically relevant contexts where machine learning is deployed, multiple platforms obtain data from the same pool of users, each of whom selects the platform that best serves them. Prior work in this setting focuses exclusively on the "l...
185. When Drafts Evolve: Speculative Decoding Meets Online Learning ​
Author: Yu-Yang Qian, Hao-Cong Wu, Yichao Fu, Hao Zhang, Peng Zhao
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.12617v2 Announce Type: replace Abstract: Speculative decoding has emerged as a widely adopted paradigm for accelerating large language model inference, where a lightweight draft model rapidly generates candidate tokens that are then verified in parallel by a larger target model. However, ...
186. Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning ​
Author: Cai Zhou, Zekai Wang, Menghua Wu, Qianyu Julie Zhu, Flora C. Shi, Chenyu Wang, Ashia Wilson, Tommi Jaakkola, Stephen Bates
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, stat.AP, stat.ML
arXiv:2604.01170v2 Announce Type: replace Abstract: While test-time scaling has enabled large language models to solve highly difficult tasks, state-of-the-art results come at exorbitant compute costs. These inefficiencies can be attributed to the miscalibration of post-trained language models, and ...
187. SODA: Semi On-Policy Black-Box Distillation for Large Language Models ​
Author: Xiwen Chen, Jingjing Wang, Wenhui Zhu, Peijie Qiu, Xuanzhao Dong, Yueyue Deng, Hejian Sang, Zhipeng Wang, Alborz Geramifard, Feng Luo
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2604.03873v5 Announce Type: replace Abstract: Black-box knowledge distillation for large language models presents a strict trade-off. Simple off-policy methods (e.g., sequence-level knowledge distillation) struggle to correct the student's inherent errors. Fully on-policy methods (e.g., Genera...
188. The Impact of Dimensionality on the Stability of Node Embeddings ​
Author: Tobias Schumacher, Simon Reichelt, Markus Strohmaier
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.08492v3 Announce Type: replace Abstract: Previous work has shown that node embedding methods can produce different representations and downstream predictions across repeated training runs, even when trained on the same data with identical hyperparameters. However, the role of embedding di...
189. Supervised Learning Has a Geometric Blind Spot ​
Author: Vishal Rajput
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2604.21395v3 Announce Type: replace Abstract: Ordinary supervised training minimises the task loss and then stops. It never pays for how far the representation moves when the input is nudged along directions that helped fit training labels---including directions that are nuisance at deployment...
190. Diffusion Operator Geometry of Feedforward Representations ​
Author: Kanishka Reddy
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, stat.ML
arXiv:2605.01107v2 Announce Type: replace Abstract: Feedforward neural networks transform data through learned representations whose geometry shapes how classes separate and relate across successive layers. We study that geometry through diffusion operators. Each feature-cloud snapshot is assigned a...
191. Realizable Bayes-Consistency for General Metric Losses ​
Author: Dan Tsir Cohen, Steve Hanneke, Aryeh Kontorovich
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, math.ST, stat.TH
arXiv:2605.03823v3 Announce Type: replace Abstract: We study strong universal Bayes-consistency in the realizable setting for learning with general metric losses, extending classical characterizations beyond $0$-$1$ classification (Bousquet et al., 2020; Hanneke et al., 2021) and real-valued regress...
192. Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination ​
Author: Jonathan Spieler, Sven Behnke
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO
arXiv:2605.04568v3 Announce Type: replace Abstract: State-of-the-art model-based Reinforcement Learning (RL) approaches either use gradient-free, population-based methods for planning, learned policy networks, or a combination of policy networks and planning. Hybrid approaches that combine Model Pre...
193. Skill Neologisms: Towards Skill-based Continual Learning ​
Author: Antonin Berthon, Nicolas Astorga, Mihaela van der Schaar
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.04970v3 Announce Type: replace Abstract: Modern LLMs show mastery over an ever-growing range of skills, as well as the ability to compose them flexibly. However, extending model capabilities to new skills in a scalable manner is an open problem: fine-tuning and parameter-efficient variant...
194. Fast Rates for Inverse Reinforcement Learning ​
Author: Andreas Schlaginhaufen, Maryam Kamgarpour
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2605.14599v2 Announce Type: replace Abstract: We establish novel structural and statistical results for entropy-regularized min-max inverse reinforcement learning (Min-Max-IRL) in finite-horizon MDPs with Borel state and action spaces. We show that maximum likelihood estimation (MLE) and Min-M...
195. CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning ​
Author: Yang Liu, Toan Nguyen, Flora D. Salim
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CV
arXiv:2605.20247v2 Announce Type: replace Abstract: Catastrophic forgetting remains a major obstacle to continual learning in large language models (LLMs) and vision--language models (VLMs). Although Mixture-of-Experts (MoE) architectures offer an efficient path to scaling, existing LoRA-based MoE c...
196. Persona-Pruner: Sculpting Lightweight Models for Role-Playing ​
Author: Jinsu Kim, Jihoon Tack, Noah Lee, Jongheon Jeong
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2606.14695v2 Announce Type: replace Abstract: Language Models (LMs) have shown remarkable potential as role-playing chatbots, delivering consistent, stylized interactions when given a specification of a character or user persona. However, applying these capabilities to real-world applications ...
197. Beyond Weights and Gradients: A Taxonomy of Federated Learning Messages ​
Author: Alvaro Javier Vargas Guerrero, Xinguang Wang, Quang Manh Doan, Guy Nagels
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.16891v2 Announce Type: replace Abstract: Federated Learning is rapidly evolving beyond the exchange of traditional model weights and gradients, yet existing definitions fail to capture the full scope of modern payloads like synthetic data and federated analytics. This paper addresses the ...
198. Right Knowledge, Wrong Answer: Characterizing Parametric Temporal Conflict in Open-Weight Language Models ​
Author: Elias Hossain, Sourav Saha, Tasfia Nuzhat Ornee, Sanjeda Sara Jennifer, Umesh Chandra Biswas, Shubhashis Roy Dipta, Rajib Rana, Niloofar Yousefi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2606.20959v3 Announce Type: replace Abstract: Language models may encode both outdated facts and their newer replacements. We introduce Parametric Temporal Conflict (PTC), where the newer fact is present and recoverable, but the default forward pass prefers the outdated one. We release a deter...
199. NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning ​
Author: Tianlin Pan, Lianyu Pang, Cheng Da, Huan Yang, Changqian Yu, Kun Gai, Wenhan Luo
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2606.27771v5 Announce Type: replace Abstract: Reinforcement learning (RL) post-training improves the reward alignment of flow-based generators, but often degrades perceptual quality in ways that are not captured by the reward proxy. We identify a simple structural signature of this drift: acro...
200. Accelerating Q-learning through Efficient Value-Sharing across Actions ​
Author: Prabhat Nagarajan, Brett Daley, Martha White, Marlos C. Machado
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.29806v2 Announce Type: replace Abstract: Action values are foundational to many control algorithms such as Q-learning. Therefore, efficient action-value learning is central to reinforcement learning (RL). However, learning them can be slow, requiring many updates to move values from their...
201. Fractal KV-Cache Archives: Lossless Symbolic Storage with In-Place Retrieval for Long-Context LLM Inference ​
Author: Vladimir Gusev
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.07144v2 Announce Type: replace Abstract: The key-value (KV) cache dominates the memory cost of long-context autoregressive inference, and a growing body of work compresses it through quantization, eviction, or offloading. We study a complementary question: once a position's KV state has b...
202. Infrared Organization and Critical Cognitive Field Formation in Transformer Dynamics ​
Author: Byung Gyu Chae
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.10923v3 Announce Type: replace Abstract: Large language models exhibit remarkable emergent behaviors, yet the physical mechanism governing their collective dynamics remains poorly understood. Cognitive Field Theory predicts that learning reorganizes the collective relaxation spectrum thro...
203. Analytic Distribution of Classifier-Free Guidance for Schedule Design ​
Author: Enze Jiang, Zheng Ma
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.19725v2 Announce Type: replace Abstract: Classifier-free guidance (CFG) is the default mechanism for conditional generation in diffusion models, but the distribution sampled by its deterministic guided dynamics is not captured by the usual product-distribution heuristic $p_0^\omega q_0^{1...
204. Variance-Preserving Orthogonal Selection (VPOS): Greedy Feature Selection via Orthogonal Deflation in PCA Loading Space ​
Author: Baran Koseoglu, Berrin Yanikoglu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.23198v3 Announce Type: replace Abstract: We present Variance-Preserving Orthogonal Selection (VPOS), an unsupervised feature-selection method that performs sequential orthogonal deflation in the variance-weighted principal component analysis (PCA) loading space $\mathbf{V}_d\mathbf{\Lambd...
205. Budget-Aware LLM Discovery via Cost-Calibrated Frontier Utility ​
Author: Yansen Zhang, Yilu Liu, Tianyu Liu, Jiamin Chen, Xiaokun Zhang, Kai Xie, Xue Liu, Yiyan Qi, Chen Ma
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26828v3 Announce Type: replace Abstract: Large language models increasingly support scientific and algorithmic discovery through inference-time search over evaluated candidates. Existing adaptive discovery controllers assign credit based only on score progress, even though prompt length, ...
206. Compliance2LoRA: Personalizable On-Demand Safety Alignment on Arbitrary Policy Subsets via Hypernetwork-Generated LoRA Adapters ​
Author: Pankayaraj Pathmanathan, Furong Huang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27594v3 Announce Type: replace Abstract: Post-training alignment in large reasoning models (LRMs) has significantly improved their adaptability to diverse safety compliance settings. However, as LRMs personalization for downstream users takes center stage, the demand for varying levels of...
207. Kohn-Sham Spectral Embedding on Sparse Graphs at the Nishimori Temperature for Image Classification ​
Author: V. S. Usatyuk, D. A. Sapozhnikov, S. I. Egorov
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.IT, math.IT
arXiv:2607.28428v3 Announce Type: replace Abstract: We propose Kohn-Sham Spectral Embedding (KSSE), an energy-based model replacing the top-layer classifier of convolutional networks with a sparse-graph spectral embedding at the Nishimori temperature of an associated Random-Bond Ising Model the spec...
208. Distill What the Student Can See: Fisher-Projected On-Policy Distillation for Vision-Language Models ​
Author: Leyan Xue, Feng Xiong, Mingjun Ma, Changqing Zhang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01263v2 Announce Type: replace Abstract: On-policy distillation (OPD) samples trajectories from the current student policy and minimizes token-level divergence between student and teacher next-token distributions at prefixes along those trajectories. This aligns the distillation states wi...
209. Output-Aware Rotation for INT2 KV-Cache Quantization ​
Author: Vincent-Daniel Yun, Woosang Lim, Minsoo Cheong, Sunwoo Lee, Murali Annavaram, Sai Praneeth Karimireddy, Sungjoo Yoo
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.02691v2 Announce Type: replace Abstract: The key-value (KV) cache has become a major memory and bandwidth bottleneck in long-context large language model inference, making ultra-low-bit quantization increasingly important. However, existing rotation-based INT2 methods optimize cache stati...
210. SAKI: Score-Aware Low-Rank Key Indexing with Random-Matrix Noise Correction for KV Retrieval ​
Author: Lin Zhang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, stat.AP
arXiv:2608.03228v2 Announce Type: replace Abstract: Existing low rank KV cache methods preserve either model weights or key variance, neither of which directly reflects the attention scores used during inference. We derive the expected attention score distortion caused by rank r key compression and ...
211. Resume Means Resume: A Machine-Checked Conformance Contract for Checkpoint, Interrupt, and Resume Semantics in Workflow Persistence Layers ​
Author: Sajjad Khan
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.LO, cs.SE
arXiv:2608.03836v2 Announce Type: replace Abstract: A framework that persists execution state so a run can be interrupted, survive a crash, and continue must decide what a resume means for effects that already fired. Five widely deployed agent workflow frameworks answer differently, none exposes a m...
212. Bi-semantic Chemical Embedder for Joint Representation Learning of SMILES and Natural Language ​
Author: David Ming Segura, Jeremy Goumaz, Joshua W. Sin, Bojana Rankovi'c, Philippe Schwaller
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.03855v2 Announce Type: replace Abstract: Transformer models have revolutionized natural language processing (NLP), and text-based molecular representations like SMILES have successfully extended these architectures to chemistry. However, domain-adaptive pre-training often causes models to...
213. Optimal Training-Time Scaling in Gradual Adaptation ​
Author: Zonghuan Xu, Krishna Harish
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.04927v2 Announce Type: replace Abstract: In gradual adaptation, how should the training time on each task change as the number of intermediate tasks increases? We study this question for overparameterized linear regression tasks that change smoothly and share a zero-loss solution. With $N...
214. Multimodal Spatiotemporal Atmospheric Data Assimilation with Latent Video Flow-matching ​
Author: Dibyajyoti Chakraborty, Romit Maulik
Published: 8/7/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, physics.ao-ph, physics.flu-dyn
arXiv:2608.05103v2 Announce Type: replace Abstract: Data assimilation (DA) uses Bayesian inference to update the state of a numerical forecast model with observed data. In this study, we propose a fundamentally different, unified approach to atmospheric data assimilation. We use latent video flow-ma...
215. Analogy as Nonparametric Bayesian Inference over Relational Systems ​
Author: Ruairidh M. Battleday, Thomas L. Griffiths
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2006.04156v2 Announce Type: replace-cross Abstract: Our inferences in the real world are rarely na"ive - we acquire experiences through our lifetime that can help us more quickly understand the structure of something new. A fundamental question in cognitive science is how we make such general...
216. Latent Utility Q-Learning for Preference-Adaptive Dynamic Treatment Regimes ​
Author: Joshua P. Zitovsky, Yating Zou, Leslie Wilson, Michael R. Kosorok
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2307.12022v3 Announce Type: replace-cross Abstract: Optimizing individualized treatment sequences for patients who weigh multiple, competing outcomes differently poses a challenge for dynamic treatment regime (DTR) methods, which typically assume a single univariate outcome. We propose Latent ...
217. A Reverse-BSDE Diffusion Sampler ​
Author: Jairon H. N. Batista, Fl'avio B. Gon\c{c}alves, Yuri F. Saporito, Rodrigo S. Targino
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR
arXiv:2505.06800v2 Announce Type: replace-cross Abstract: Diffusion-based generative models have renewed interest in stochastic differential equation methods for sampling from complex distributions. We study a setting in which the target density is known only up to a normalizing constant and reformu...
218. MoCA: Multi-modal Cross-masked Autoencoder for Time Series in Digital Health ​
Author: Howon Ryu, Yuliang Chen, Yacun Wang, Andrea Z. LaCroix, Chongzhi Di, Loki Natarajan, Yu Wang, Jingjing Zou
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP
arXiv:2506.02260v5 Announce Type: replace-cross Abstract: Wearable devices enable continuous multi-modal physiological and behavioral monitoring, yet analysis of these data streams faces fundamental challenges including the lack of gold-standard labels and incomplete sensor data. While self-supervis...
219. CPC-CMS: Cognitive Pairwise Comparison Classification Model Selection Framework for Document-level Sentiment Analysis ​
Author: Jianfei Li, Kevin Kam Fung Yuen
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2507.14022v2 Announce Type: replace-cross Abstract: This study proposes the Cognitive Pairwise Comparison Classification Model Selection (CPC-CMS) framework for document-level sentiment analysis. The CPC, based on expert knowledge judgment, is used to calculate the weights of evaluation criter...
220. Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet ​
Author: James Xu Zhao, Bryan Hooi, See-Kiong Ng
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2509.06861v3 Announce Type: replace-cross Abstract: Test-time scaling increases inference-time computation through longer reasoning chains and has shown strong performance gains across many domains. However, frontier models still suffer from factuality hallucinations, raising the question of w...
221. Dynamic Object Masks as Goal Representations for Visual Goal-Conditioned Reinforcement Learning ​
Author: Fahim Shahriar, Cheryl Wang, Alireza Azimi, Gautham Vasan, Hany Hamed, Abhishek Naik, A. Rupam Mahmood, Colin Bellinger
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2510.06277v2 Announce Type: replace-cross Abstract: Goal-conditioned reinforcement learning (GCRL) offers a unified way to pursue diverse tasks, yet most existing methods rely on state- or position-based goal representations that are unavailable in real-world robotics. Robots operating in ware...
222. Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts ​
Author: Emanuele Marconato, Samuele Bortolotti, Emile van Krieken, Paolo Morettin, Elena Umili, Antonio Vergari, Efthymia Tsamoura, Andrea Passerini, Stefano Teso
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2510.14538v3 Announce Type: replace-cross Abstract: Neuro-symbolic (NeSy) AI aims to develop deep neural networks whose predictions comply with prior knowledge encoding, e.g. safety or structural constraints. As such, it represents one of the most promising avenues for reliable and trustworthy...
223. MermaidSeqBench: An Evaluation Benchmark for NL-to-Mermaid Sequence Diagram Generation ​
Author: Basel Shbita, Farhan Ahmed, Chad DeLuca
Published: 8/7/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2511.14967v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown great promise in generating structured diagrams from natural language descriptions, particularly Mermaid sequence diagrams for software engineering. However, the lack of existing benchmarks to assess th...
224. A note on conditional PAC-efficient reasoning in large language model routing ​
Author: Hao Zeng, Bingyi Jing
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, math.ST, stat.TH
arXiv:2512.03057v2 Announce Type: replace-cross Abstract: We study distribution-free risk control for model routing, motivated by large language model reasoning. We formalize pointwise conditional efficiency under a probably approximately correct guarantee and show that it forces a nearly impossible...
225. Deterministic World Models for Closed-loop Reachability Analysis of End-to-End Vision-based Control ​
Author: Yuang Geng, Zhongzheng Zhang, Chengzhen Jiang, Yanru Li, Xinyang Wang, Zhuoyang Zhou, Hoang-Dung Tran, Ivan Ruchkin
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2512.08991v3 Announce Type: replace-cross Abstract: End-to-end image controllers that map raw camera frames directly to control actions are increasingly deployed in safety-critical systems. However, formally verifying their closed-loop behavior remains an open challenge because cameras produce...
226. Perfect reconstruction of sparse signals using nonconvexity control and one-step RSB message passing ​
Author: Xiaosi Gu, Ayaka Sakata, Tomoyuki Obuchi
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.dis-nn, cs.LG
arXiv:2512.17426v2 Announce Type: replace-cross Abstract: We consider sparse signal reconstruction via minimization of the smoothly clipped absolute deviation (SCAD) penalty, and develop one-step replica-symmetry-breaking (1RSB) extensions of approximate message passing (AMP), termed 1RSB-AMP. Start...
227. EqDeepRx: Learning a Scalable and Interference Mitigating MIMO Receiver ​
Author: Mikko Honkala, Dani Korpi, Elias Raninen, Janne M. J. Huttunen
Published: 8/7/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2602.11834v2 Announce Type: replace-cross Abstract: While machine learning (ML)-based receiver algorithms have received a great deal of attention in the recent literature, they often suffer from poor scaling with increasing spatial multiplexing order and lack of explainability and generalizati...
228. STATe-of-Thoughts: Structured Action Templates for Tree-of-Thoughts ​
Author: Zachary Bamberger, Till R. Saenger, Gilad Morad, Ofra Amir, Brandon M. Stewart, Amir Feder
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2602.14265v3 Announce Type: replace-cross Abstract: Inference-Time-Compute (ITC) methods like Best-of-$n$ and Tree-of-Thoughts are meant to produce output candidates that are both high-quality and diverse, but their use of high-temperature sampling often fails to achieve meaningful output dive...
229. Dual-space posterior sampling for Bayesian inference in constrained inverse problems ​
Author: Ali Siahkoohi, Kamal Aghazade, Ali Gholami
Published: 8/7/2026, 4:00:00 AM
Categories: physics.geo-ph, cs.LG, stat.ML
arXiv:2603.00393v2 Announce Type: replace-cross Abstract: Inverse problems constrained by partial differential equations are often ill-conditioned due to noisy, incomplete data or inherent non-uniqueness. A prominent example is full waveform inversion (FWI), which estimates Earth's subsurface proper...
230. Clinician input steers AI toward accurate and harmful recommendations ​
Author: Ivan Lopez, Selin S. Everett, Bryan J. Bunning, April S. Liang, Dong Han Yao, Shivam C. Vedak, Kameron C. Black, Sophie Ostmeier, Stephen P. Ma, Emily Alsentzer, Jonathan H. Chen, Akshay S. Chaudhari, Eric Horvitz
Published: 8/7/2026, 4:00:00 AM
Categories: cs.HC, cs.LG
arXiv:2603.14158v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are entering clinical workflows, yet evaluations rarely assess how clinician reasoning shapes model behavior during clinical interactions. Using 61 curated NEJM Case Records, we tested how expert or misleading cli...
231. Communication-Aware Multi-Agent Reinforcement Learning for Decentralized Cooperative UAV Deployment ​
Author: Enguang Fan, Yifan Chen, Zihan Shan, Klara Nahrstedt, Matthew Caesar, Jae Kim
Published: 8/7/2026, 4:00:00 AM
Categories: cs.MA, cs.LG, cs.NI
arXiv:2603.16141v2 Announce Type: replace-cross Abstract: Autonomous Unmanned Aerial Vehicle (UAV) swarms are increasingly used as rapidly deployable aerial relays and sensing platforms, yet practical deployments must operate under partial observability and intermittent peer-to-peer connectivity. We...
232. NavTrust: Benchmarking Trustworthiness for Embodied Navigation ​
Author: Huaide Jiang, Yash Chaudhary, Yuping Wang, Zehao Wang, Raghav Sharma, Manan Mehta, Yang Zhou, Lichao Sun, Zhiwen Fan, Zhengzhong Tu, Jiachen Li
Published: 8/7/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG, cs.SY, eess.SY
arXiv:2603.19229v2 Announce Type: replace-cross Abstract: There are two major categories of embodied navigation: Vision-Language Navigation (VLN), where agents navigate by following natural language instructions; and Object-Goal Navigation (OGN), where agents navigate to a specified target object. H...
233. {\lambda}Split: Self-Supervised Content-Aware Spectral Unmixing for Fluorescence Microscopy ​
Author: Federico Carrara, Talley Lambert, Mehdi Seifi, Florian Jug
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2603.23647v3 Announce Type: replace-cross Abstract: In fluorescence microscopy, spectral unmixing aims to recover individual fluorophore concentrations from spectral images that capture mixed fluorophore emissions. Since classical methods operate pixel-wise and rely on least-squares fitting, t...
234. Assessing the Role of Intersection Proximity in Pedestrian Crashes: Insights from Data Mining Approach ​
Author: Ahmed Hossain, Xiaoduan Sun
Published: 8/7/2026, 4:00:00 AM
Categories: physics.soc-ph, cs.LG
arXiv:2604.28065v2 Announce Type: replace-cross Abstract: Although intersections are the most complex parts of the roadway network, pedestrian crashes at non-intersection locations are disproportionately frequent, highlighting a serious traffic safety concern. This study investigates non-intersectio...
235. MoDAl: Self-Supervised Neural Modality Discovery via Decorrelation for Speech Neuroprosthesis ​
Author: Yuanhao Chen, Peter Chin
Published: 8/7/2026, 4:00:00 AM
Categories: q-bio.NC, cs.CL, cs.HC, cs.LG, eess.AS
arXiv:2605.00025v3 Announce Type: replace-cross Abstract: Speech neuroprosthesis systems decode intended speech from neural activity in the absence of audible output, offering a path to restoring communication for individuals with speech-impairing conditions. Current approaches decode predominantly ...
236. The Impossibility Triangle of Long-Context Modeling ​
Author: Yan Zhou
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2605.05066v2 Announce Type: replace-cross Abstract: We identify and prove a fundamental trade-off governing long-sequence models: no model can simultaneously achieve (i) per-step computation independent of sequence length (Efficiency), (ii) state size independent of sequence length (Compactnes...
237. Positive-Data Learning of Fixed-Observation Linear MCFGs from Working Binary Presentations ​
Author: Takayuki Kuriyama
Published: 8/7/2026, 4:00:00 AM
Categories: cs.FL, cs.LG
arXiv:2605.11644v2 Announce Type: replace-cross Abstract: We study positive-data learning of languages admitting reduced working binary linear nondeleting multiple context-free grammar presentations of bounded fan-out. The learner is supplied with a fixed explicit finite monoid homomorphism (h:\Sigm...
238. Phylogenetic Tree Inference with Tropical Axial Attention ​
Author: Chris Teska, Kurt Pasque, Ruriko Yoshida, Baran Hashemi
Published: 8/7/2026, 4:00:00 AM
Categories: q-bio.PE, cs.LG
arXiv:2605.13894v2 Announce Type: replace-cross Abstract: In this work, we introduce a Tropical Axial Attention neural reasoning architecture that replaces vanilla softmax dot-product attention with max-plus operators, inducing a piecewise-linear structure aligned with dynamic programming formulatio...
239. Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift ​
Author: Qinwu Xu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.DB, cs.LG
arXiv:2605.16411v2 Announce Type: replace-cross Abstract: Hallucination remains a fundamental challenge in vision-language models (VLMs), where autoregressive generation may produce linguistically plausible yet physically inconsistent or visually ungrounded responses due to likelihood maximization u...
240. Beyond Adoption Intention How Trust in Augmented Analytics Relates to Perceived Decision Quality Among Non-Technical BI Users ​
Author: Thuy Pham Thi Phuong, Hieu Vu Le Trung, Ha Nguyen Manh, Nhi Tran Pham Yen, Lan Hoang Thi
Published: 8/7/2026, 4:00:00 AM
Categories: cs.HC, cs.CY, cs.LG
arXiv:2605.20198v3 Announce Type: replace-cross Abstract: Augmented analytics has transformed how Business Intelligence (BI) systems support decision-making, shifting non-technical managers from manual analysis toward dependence on automated insights. Current BI research often overlooks the cognitiv...
241. Physics-Guided Concentration Inference from Resistance Transients in a Mixed-Phase SnO-SnO$_2$ Carbon Monoxide Sensor with p-n Switching ​
Author: Sani Biswas, Preetam Singh, Amit Kumar Gangwar
Published: 8/7/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, physics.app-ph
arXiv:2605.23971v2 Announce Type: replace-cross Abstract: This work presents a physics-guided machine-learning framework for carbon monoxide concentration inference from experimentally measured resistance transients of a mixed-phase SnO-SnO$_2$ material gas sensor exhibiting temperature-dependent p-...
242. What Drives Test-Time Adaptation for CLIP? A Controlled Empirical Study from an Update Perspective ​
Author: Jiazhen Huang, Xiao Chen, Zhiming Liu, Yaru Sun, Jingyan Jiang, Zhi Wang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2606.14299v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) such as CLIP have become a standard backbone for open-vocabulary recognition, yet their zero-shot predictions remain vulnerable to distribution shifts encountered at deployment. Test-Time Adaptation (TTA) has rec...
243. ClayBuddy: A Framework, Evaluation, & Mitigation of Coding Agent Failures ​
Author: Kenneth Ge, Andre Assis
Published: 8/7/2026, 4:00:00 AM
Categories: cs.SE, cs.LG
arXiv:2606.19380v4 Announce Type: replace-cross Abstract: Software engineering and deployment are increasingly delegated to AI coding agents. The scale of their adoption is surfacing rare, but highly destructive, failure modes. In this paper, we study these failure modes as stemming from three disti...
244. Time Series Classification through Diffeomorphic Time Warping (DiffTW) ​
Author: Vicky Geneva Haney, Kamel Lahouel, Victor Rielly, Bruno M. Jedynak
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2606.23472v2 Announce Type: replace-cross Abstract: Time series classification involves learning a mapping from a continuous, temporally ordered sequence of real-valued observations to discrete response variables, like class labels. This task is fundamental in domains, including health monitor...
245. TESSERA v2: Scaling Pixel-wise Earth Foundation Models ​
Author: Zhengpeng Feng, Sadiq Jaffer, Ira Shokar, Jovana Knezevic, James Ball, Pedro Sousa, Mark Elvers, Madeline Lisaius, Clement Atzberger, Robin Young, Aneesh Naik, Niall Robinson, David Coomes, Anil Madhavapeddy, Srinivasan Keshav
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.03949v2 Announce Type: replace-cross Abstract: Pixel-wise Earth-observation (EO) foundation models are now achieving state-of-the-art performance via generated spatial embeddings. However, how these models scale and how best to spend a pretraining budget remain poorly understood. We prese...
246. Benchmark Evaluation of Federated Learning on Multi-organ Images ​
Author: Junbin Mao, Xu Tian, Jianchun Zhu, Ludi Li, Jin Liu
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.08219v2 Announce Type: replace-cross Abstract: The privacy requirements of medical data and its substantial variations across organs and modalities hinder the clinical implementation of medical AI. Federated learning (FL) is a feasible approach to overcome these challenges. Due to the con...
247. Recti-Q: Feature-Space Rectification for Out-of-Distribution-Robust Quantized Perception in Edge Robotics ​
Author: Hamidreza Yaghoubi Araghi, Parastoo Pilevar, Ming C. Lin
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO
arXiv:2607.18540v2 Announce Type: replace-cross Abstract: Robotic perception pipelines increasingly rely on large vision backbones deployed on SWaP-constrained edge platforms, making post-training quantization (PTQ) attractive for real-time inference. However, while PTQ often preserves clean in-dist...
248. H+ Embedding: Harmonizing Global and Token-Level Retrieval with Context-Dependent Phrases ​
Author: Shusen Zhang, Junyi Hu, Ye Feng, Ziteng Wang, Zhaoyuan Pan, Guosheng Dong, Xiaojun Yuan, Jiangshou Hong, Xiangzhi Wang
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.00065v2 Announce Type: replace-cross Abstract: Terminology-intensive retrieval, especially in medical settings, depends on preserving multi-word entities, abbreviations, numerical constraints, and compositional concepts. However, existing representations lie at two extremes: single-vector...
249. Rapid Embodiment Adaptation for Quadrupedal Locomotion ​
Author: Dichen Li, Bo Ai, Nico Bohlinger, Jan Peters, Hao Su, Henrik I. Christensen
Published: 8/7/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2608.01506v2 Announce Type: replace-cross Abstract: Humans readily adapt their movements as their bodies change through aging, injury, or load carrying, but learning-based robot policies often break when hardware properties shift. We introduce an online embodiment adaptation framework for quad...
250. Field Aware Agent Skill Retrieval ​
Author: Paimon Goulart, Liang Wu, Kelly Wan, Evangelos E. Papalexakis, Liangjie Hong
Published: 8/7/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.02880v2 Announce Type: replace-cross Abstract: As lifelong learning agents accumulate lifelong growing skill banks, retrieving the correct skill becomes an increasingly important bottleneck. Most current skill retrieval methods treat each skill as one flat document by concatenating fields...
251. Stochastic Saddle Avoidance Beyond Unit Excitation and Smoothness: A Pathwise Lyapunov-Perron Framework ​
Author: Junwen Qiu, Bohao Ma, Andre Milzarek, Junyu Zhang
Published: 8/7/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DS, stat.ML
arXiv:2608.03001v2 Announce Type: replace-cross Abstract: Unit excitation (UE) is a common assumption in stochastic saddle avoidance: the stochastic error must have a uniformly positive component along every direction, in expectation. This condition gives a direct way to rule out convergence to stri...
252. Surrogate Substitution Preserves PHI Detectability: A Multi-Detector Equivalence Study ​
Author: Qiming Bao, Sherry J. H. Feng, Kim Chester Eugenio, Meng Fon
Published: 8/7/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.03172v2 Announce Type: replace-cross Abstract: Structure-preserving de-identification replaces protected health information (PHI) with realistic same-type surrogates -- "Anna S." becomes "Maria S.", not [NAME] -- so that clinical text stays fluent and downstream tools keep working. But th...
253. SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs ​
Author: Kejian Zhu, Zhuoran Jin, Shangqing Tu, Hongbang Yuan, Yushi Bai, Kang Liu, Juanzi Li, Jun Zhao
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.03573v2 Announce Type: replace-cross Abstract: Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) exhibit fundamentally different behaviors in enhancing multi-task reasoning for large language models (LLMs). Our preliminary experiments revealed a phenomenon: SFT suffers from sev...
254. Accelerating Dynamic Graph Clustering on GPU Architectures with cuGraph ​
Author: Nelson Aloysio Reis de Almeida Passos, Emanuele Carlini, Salvatore Trani
Published: 8/7/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2608.03695v2 Announce Type: replace-cross Abstract: This work addresses community detection in temporal networks through GPU-accelerated extensions of spectral clustering and modularity-based algorithms originally designed for static graphs. Built on the NVIDIA RAPIDS ecosystem, the framework ...
255. Monsoon Mayhem to Market Waves: Forecasting Fisheries Resilience in Sri Lanka ​
Author: Ruzaini Ahmed, Yohan Jayasinghe, Tharumini Gamage, Ifaz Ikram, Hasini Lawanya, Nirasha Munasinghe, Patalee Narasinghe, Nisansa de Silva, Sandareka Wickramanayake
Published: 8/7/2026, 4:00:00 AM
Categories: econ.GN, cs.LG, q-fin.EC
arXiv:2608.04023v2 Announce Type: replace-cross Abstract: Sri Lanka's fisheries sector is important for jobs and food supply. Between 2019 and 2025, it faced several major problems at the same time, and how these events together affected fish production and prices is still not well understood. This ...
256. A Counterexample to Fourier Alignment in Single-Neuron Modular Addition ​
Author: Gautam Neelakantan Memana
Published: 8/7/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.NE
arXiv:2608.04451v2 Announce Type: replace-cross Abstract: We give a negative solution to MAIS-O60. We first construct an example in which an initially active ReLU neuron becomes completely inactive in finite time and thereafter remains frozen at a limit whose Fourier energy is equally distributed am...
257. EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive Decision Tasks ​
Author: Pau Arnal, Khaled Denfir, Danylo Smahliuk, Amrut Avhad, Marcus A. Castro
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.04549v2 Announce Type: replace-cross Abstract: Frontier LLMs are increasingly put to use on open-ended complex questions, different in nature from the ones they are typically evaluated on. We dedicate more than 4,000 human expert hours to evaluate a selection of six frontier LLMs on a mem...
258. Nonparametric Goodness-of-fit Testing under Covariate Shift ​
Author: Zhen Hou, Dong Xia
Published: 8/7/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, math.ST, stat.ML, stat.TH
arXiv:2608.04860v2 Announce Type: replace-cross Abstract: This paper develops procedures for nonparametric goodness-of-fit testing under covariate shift, where labelled data are drawn from a source population but goodness-of-fit is evaluated for a target population. The distribution mismatch is quan...
259. A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination ​
Author: Wenxiao Zhao, Dong Liu, Kaiyi Xu, Feng Liu, Zhen Zhao, Fei Ben, Shu Wang, Wenhao Li, Ying Nian Wu, Fenghua Ling, Haobo Li, Lei Bai
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.04872v2 Announce Type: replace-cross Abstract: Symbolic regression aims to discover closed-form equations from data, but existing LLM-guided methods often rely on a unified proposal loop that compresses heterogeneous search failures into a scalar score and a single prompt. We propose A-SR...
260. Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes ​
Author: Junlin Han, Shengbang Tong, David Fan, Minghao Chen, Philip Torr, Filippos Kokkinos, Mike Lewis
Published: 8/7/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.MM
arXiv:2608.05000v2 Announce Type: replace-cross Abstract: Vision offers a critical axis for advancing foundation models, driving a shift towards natively unified multimodal pretraining. Despite this momentum, the design space and the fundamental mechanisms of how modalities interact during unified t...