arXiv cs.LG - 2026-08-04 ​
565 items collected.
1. Uncertainty-Aware Simulation-Based Inference for Operations Research with Large Language Models ​
Author: Liang Guo, Lin Shaochong, Shen Zuo-Jun Max, Zhang Kun
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.00019v1 Announce Type: new Abstract: Deploying large language models (LLMs) for operations research (OR) tasks remains challenging because correctness depends on a coherent modeling process, not merely a correct final answer. Standard autoregressive generation operates on a myopic policy,...
2. Learning Compositional Meta-Routing for Agentic Workflows: An Executable Benchmark ​
Author: Natan Vidra, Alina Kapanova, Arun Kanhai, Spurthi Setty
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00106v1 Announce Type: new Abstract: Agentic systems must decide not only what answer to produce, but which reasoning and execution operations should precede it. A controller may answer directly, decompose a request, retrieve evidence, execute code, delegate to a specialist, or verify an ...
3. MetaRoute-Bench: Evaluating Meta-Decision Policies for Agentic Workflow Routing ​
Author: Natan Vidra, Alina Kapanova, Arun Kanhai, Spurthi Setty
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.00107v1 Announce Type: new Abstract: Agentic systems must repeatedly decide whether to answer directly, decompose a task, invoke a tool, execute code, delegate to a specialist, verify an intermediate result, or recover from failure. These meta-decisions affect not only task success but al...
4. Progressive$^2$: A Teacher-Student Progressive Co-Evolving Knowledge Distillation Method for Substantial Model Compression ​
Author: Tiancong Cheng, Ying Zhang, Zhiwen Yu, Yifang Yin, Bin Guo
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00129v1 Announce Type: new Abstract: Knowledge distillation (KD) is a widely utilized technique for transferring knowledge from a large model (the teacher) to a smaller model (the student). Owing to its flexibility and broad applicability, KD has been extensively applied in the compressio...
5. Rethinking Pretraining for Specialized Design Data: Evidence from the JONES-19 Cultural Design Dataset ​
Author: Alexandros Haridis, Charles Zhou
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.00135v1 Announce Type: new Abstract: Design and architectural archives encode expert human knowledge in graphical formats, providing a critical testbed for design-inspired Machine Learning (ML) challenges absent with typical computer vision benchmarks. Building on JONES-19, a small-size i...
6. Leak It: A Probabilistic Approach to Training-Data Extraction from Black-Box Language Models ​
Author: Victor Maricato
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CR
arXiv:2608.00144v1 Announce Type: new Abstract: Membership inference (MIA) on language models is usually summarised by an aggregate ROC-AUC, but such evaluations are confounded: model-free blind baselines separate members from non-members from surface text alone. We study black-box, sampling-based t...
7. Response Magnitude as a Dominant Signal for Held-Out CRISPRi Perturbation Effect Prediction ​
Author: Mehrdad Shoeibi, Niloofar Yousefi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.00152v1 Announce Type: new Abstract: Predicting the magnitude of a CRISPRi perturbation's transcriptomic effect on held-out target genes is an important open problem in single-cell biology. Recent work has documented that simple baselines often match or exceed deep perturbation predictors...
8. Inference-Time Policy Alignment for Fair Reinforcement Learning ​
Author: Umer Siddique, Peilang Li, Conor Wallace, Yongcan Cao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.00175v1 Announce Type: new Abstract: Deep reinforcement learning (RL) agents achieve strong performance by optimizing scalar reward functions. However, once deployed, the policies of these RL agents are often rigid and costly to adapt to new performance criteria. For instance, an agent tr...
9. AutoCause: A Python framework that automates expert decisions in environmental time-series causal discovery ​
Author: Marco Ruiz, Miguel Arana-Catania, David R. Ardila, Rodrigo Ventura
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00198v1 Announce Type: new Abstract: Environmental time-series causal discovery requires expert decisions about method choice, conditional-independence tests, lag horizons, sample-size adequacy, multiple-testing control, and evidence interpretation. Applied inconsistently across datasets,...
10. A Physics-Chemistry-Informed Neural Network (PCINN) for Real-Time Spatial-ALD Coverage Prediction and Reliable Kinetics Inversion ​
Author: Ning Hu, Chang Liu, Yunlei Jiang, Yuan Dong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2608.00212v1 Announce Type: new Abstract: Spatial atomic layer deposition (SALD) is a leading atmospheric-pressure, high-throughput route to industrial ALD, but design and control are limited by the cost of predicting surface coverage: high-fidelity CFD is far too slow for operating-window sca...
11. Verifier-Induced Support Reshaping in On-Policy Optimization ​
Author: Shaohang Wei, Zikun Su, Feifan Song, Wen Luo, Wei Li, Guangyue Peng, Houfeng Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.00220v1 Announce Type: new Abstract: We show that on-policy reinforcement learning with verifiable rewards (RLVR) can improve the current objective while making successful behaviors for later objectives too rare to sample and reinforce. We call this verifier-induced support reshaping and ...
12. Similarity-Aware Machine Unlearning ​
Author: Madhavan Citalamangalam Kumaran, Midhun Parakkal Unni, Vicky Kouni, Haripriya Harikumar
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00246v1 Announce Type: new Abstract: Machine unlearning removes the influence of user-specified training examples from a trained model, avoiding the need to retrain it from scratch. Localization-based methods improve unlearning efficiency by identifying a subset of influential model param...
13. Stabilized Best-of-$K$ Training for Neural Combinatorial Optimization ​
Author: Melveena Jolly, Midhun Xavier
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00296v1 Announce Type: new Abstract: Leader Reward modifies POMO training to emphasize the best trajectory produced by repeated inference. We test a narrow extension: replace its binary leader/non-leader distinction with a stabilized rank signal indexed by a sampling budget $K$. With the ...
14. Abstention as an Action Can Kill Both the Reward Gradient and the KL Anchor: Collapse Law and Repair for Error-Penalized Reinforcement Learning ​
Author: Xujun Che, Yuchen Yuan, Weida Zhao, Chenyang Yu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.00301v1 Announce Type: new Abstract: Error-penalized scoring rules ($+1$ for a correct answer, $-\lambda$ for a wrong one, $0$ for abstaining) are increasingly prescribed against hallucination: a rational agent facing such a rule answers exactly when its correctness probability exceeds Ch...
15. Agentic Bayesian Optimization through Surrogate-Augmented Autoresearch ​
Author: Paul Brunzema, Louis Tiao, Nhat Le, Kevin De Angeli, Yao Xuan, Djordje Gligorijevic
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.00316v1 Announce Type: new Abstract: Bayesian optimization (BO) has become the standard tool for sample-efficient optimization and owes its efficiency to uncertainty-aware search driven by generic statistical priors. Richer domain priors can improve BO in principle, but encoding them thro...
16. Neural operator learning for collision-aware trajectory planning of spacecraft swarms ​
Author: Sidhdharth D. Sikka, Suyi Gao, Zehui Lu, Rongjie Lai, Shaoshuai Mou
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.MA, cs.SY, eess.SY
arXiv:2608.00320v1 Announce Type: new Abstract: Autonomous spacecraft swarms must plan fuel-efficient, collision-free maneuvers in increasingly congested orbits, yet classical trajectory optimization scales poorly as pairwise safety constraints multiply with swarm size, and learning-based planners r...
17. Ensemble of Unsupervised Deep Learning for Clustering Imbalanced Tabular Data ​
Author: Pulock Das, Yina Hou, Md. Kamrozzaman Bhuiyan, Manar D. Samad
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00346v1 Announce Type: new Abstract: Data imbalance poses a major challenge in supervised classification, where the majority-class bias contributes to false negatives and overestimates classification accuracy. Unsupervised deep clustering can be immune to class imbalance because represent...
18. Modeling Unknown Nonlocal PDE Systems via Flow Map Learning ​
Author: Zhongshu Xu, Ying Li, Yanzhi Zhang, Dongbin Xiu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.DS, math.NA
arXiv:2608.00400v1 Announce Type: new Abstract: Nonlocal partial differential equations arise in many applications but are often difficult to model and learn because of the presence of nonlocal operators. We present a flow-map learning (FML) framework for modeling unknown nonlocal PDEs directly from...
19. DSETA: A Dual-Stage Continual Learning Framework for Travel Time Prediction in Dynamic Traffic Environments ​
Author: Yanming Lyu, Yue Cheng, Lingkun Li, Ruipeng Gao, Xinyue Liu, Hui Gao, Qiang Ni
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.00402v1 Announce Type: new Abstract: Estimated Time of Arrival (ETA) prediction is a core component of intelligent transportation systems. As traffic congestion patterns become increasingly dynamic in large cities, maintaining high prediction accuracy poses a major challenge for ride-hail...
20. Unleashing the Potential of Large Language Models: A Blueprint for Real-Time, Enterprise-Ready Deployments ​
Author: Muhammad Faizan Raza (Luna), Shuo (Luna), Yang, Satish Mahadevan Srinivasan, Joanna F. DeFranco
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.IR
arXiv:2608.00419v1 Announce Type: new Abstract: Large language models deployed in real-time, regulated settings face knowledge staleness, catastrophic forgetting, hallucination, and weak feedback loops. We present a unified, pattern-driven LLMOps architecture integrating real-time data ingestion, co...
21. HP-JEPA: Hierarchical Partitioning for Multi-Resolution Graph Joint-Embedding Predictive Learning ​
Author: Ruichen Xu, Jingxiang Qu, Wenhan Gao, Jiaxing Zhang, Linsey Pang, Ravid Shwartz-Ziv, Yann LeCun, Yuefan Deng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00491v1 Announce Type: new Abstract: Graph self-supervised learning aims to learn transferable representations from large-scale unlabeled graph data. Joint-embedding predictive architectures (JEPAs) avoid explicit negative-pair construction and raw-input reconstruction by predicting maske...
22. Agentic Graph Token Reasoning ​
Author: Zhuoyi Peng, Yi Yang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00542v1 Announce Type: new Abstract: Graphs model relational data throughout science and industry, from citation networks to product co-purchase graphs. Because the nodes of many such graphs carry rich text, a growing line of work applies large language models (LLMs) to graph analysis. Th...
23. Crushing the Evidence: A Dual-Penalty Evasion Framework for Fooling White-Box Explainable AI Auditors ​
Author: Niraj Kumar, Harsh Kasyap
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.00566v1 Announce Type: new Abstract: Post-hoc model explainers such as LIME, SHAP, and Integrated Gradients are widely deployed to audit models in high-stakes sensitive domains, including finance, healthcare, and social welfare. This ensures the model's transparency and acceptability. How...
24. Fairness Auditing: Lower Bounds on Company Manipulation ​
Author: Rachit Verma, Padala Manisha, Sujit Gujar
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.00568v1 Announce Type: new Abstract: Fairness audits are increasingly mandated in high-stakes applications such as hiring, lending, and automated decision-making. Recent work has established fundamental impossibility results for black-box fairness auditing, showing that sufficiently expre...
25. CoSynFlow: Conformal Symplectic Neural Flows for Cross-System Prediction of Dissipative Hamiltonian Dynamics ​
Author: Baige Xu, Takaharu Yaguchi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00571v1 Announce Type: new Abstract: Learning solution operators for differential equations is a central problem in scientific machine learning. However, many neural operator methods optimize prediction accuracy without explicitly enforcing the geometric structure of the dynamics. Structu...
26. From field-scale to large-scale spectral libraries: Tabular foundation models in soil spectroscopy ​
Author: Viacheslav Barkov, Jonas Schmidinger, Robin Gebbers, Martin Atzmueller
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00608v1 Announce Type: new Abstract: Visible and near-infrared (vis-NIR) and mid-infrared (MIR) spectroscopy enable rapid, cost-effective prediction of soil properties. Yet, translating high-dimensional, highly collinear spectra into accurate soil property predictions remains challenging,...
27. RHEA: Reliability-Harmonized Reconstruction and Assignment for Robust Multimodal-Attributed Graph Clustering ​
Author: Yinlin Zhu, Di Wu, Ziyu Han, Zekai Chenm, Wang Luo, Miao Hu, Guocong Quan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00621v1 Announce Type: new Abstract: Multimodal-attributed graphs (MAGs), whose nodes carry heterogeneous attributes such as text and images over a relational structure, have become a fundamental substrate for label-free entity grouping tasks, including community discovery and product seg...
28. Towards Effective Federated Multimodal Graph Learning via Navigating Multifaceted Heterogeneity ​
Author: Yinlin Zhu, Di Wu, Yi Zhang, Xunkai Li, Wang Luo, Wei-Jin Huang, Miao Hu, Guocong Quan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00623v1 Announce Type: new Abstract: Multimodal-attributed graphs (MAGs), where nodes carry heterogeneous semantic content across multiple modalities while edges encode relational dependencies, have been widely adopted across diverse domains. Federated multimodal graph learning (FMGL) ext...
29. Relative Parameter Importance in Task-Agnostic Replay-Free Continual Learning ​
Author: Malavika Suresh, Ikechukwu Nkisi-Orji, Nirmalie Wiratunga
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00630v1 Announce Type: new Abstract: Achieving continual learning (CL) with deep neural networks requires balancing stability and plasticity while enabling knowledge transfer. In this work, we focus on offline learning algorithms under the constraints: (I) no access to training data from ...
30. Learning the Pareto Frontier of Predictive Models under Distribution Shift ​
Author: Yiming Dong, Jiwei Zhao, Yang Young Lu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.00632v1 Announce Type: new Abstract: Modern machine learning pipelines increasingly rely on reusing pretrained and foundation models across downstream tasks. These pretrained models can differ not only in performance but also in how they can be used: some only provide black-box prediction...
31. Mitigating Backdoors via Decoy Shortcuts and Knowledge Decoupling ​
Author: Zixuan Zhu, Rui Wang, Lihua Jing, Jinwen Zhong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.00732v1 Announce Type: new Abstract: Backdoor attacks pose a serious threat to deep neural networks, especially when training relies on third-party data, allowing adversaries to inject malicious behaviors through data poisoning. In this work, we reveal that backdoor behaviors tend to be a...
32. An Embedded RISC-V Evaluation of Kolmogorov--Arnold Networks in Hard-Constrained Recurrent Physics-Informed Models ​
Author: Enzo Nicolas Spotorno, Josafat Leal Filho
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.PF
arXiv:2608.00737v1 Announce Type: new Abstract: Hard-constrained recurrent physics-informed networks (HRPINNs) embed known dynamics inside a recurrent numerical integrator and restrict a neural branch to learning only the residual dynamics that the first-principles model does not capture. Kolmogorov...
33. Generic Vision and Cross-Attention for Reaction Yield Prediction ​
Author: Qiwei Han, Chi Zhou
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph
arXiv:2608.00776v1 Announce Type: new Abstract: Traditional reaction yield prediction is constrained by 1D quantum descriptors that lack explicit spatial information. To address this gap, a dual-modal Vision Cross-Attention architecture is proposed, fusing tabular physical-organic data with 2D molec...
34. Paris as a 15-Minute City: An Explainable AI Perspective ​
Author: Andr'as J. Moln'ar, Csaba I. Sidl'o, Rita R'onai, Domonkos R'ozsay
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00815v2 Announce Type: new Abstract: The 15-minute city promotes access to everyday services within a short walk or bicycle ride, but its relationship with observed mobility remains difficult to quantify. We investigate this relationship in the Paris metropolitan area using mobility traje...
35. AdvPlan-Bench: Adversarial Evaluation of Structured Plan-Generation Agents ​
Author: Alina Kapanova, Arun Kanhai, Natan Vidra, Spurthi Setty
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00832v1 Announce Type: new Abstract: Structured plan-generation agents are often evaluated as if a plan has quality in isolation, yet many realistic planning tasks require asking how a candidate behaves when another agent can search for responses. We introduce AdvPlan-Bench, an offline be...
36. Deep Learning CNN and Recurrence Analysis for Alpha Gamma EEG Biomarkers in Fragile X Syndrome ​
Author: Zag ElSayed, Payton Siekierski, Jack Yanchen Liu, Ernest Pedapati
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, physics.data-an, q-bio.NC
arXiv:2608.00835v1 Announce Type: new Abstract: Fragile X Syndrome (FXS) is a neurodevelopmental disorder caused by reduced expression of fragile X mental retardation protein (FMRP), leading to disrupted synaptic plasticity, cortical hyperexcitability, and impaired network synchronization. Electroen...
37. Nonlinear Laplacians Improve Signed-Directed Graph Learning ​
Author: Ali Parviz, Yuichi Yoshida
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00836v1 Announce Type: new Abstract: While signed-directed graphs have been studied using linear Laplacians in the design of graph neural networks, relatively little research has focused on developing non-linear Laplacian operators for such networks. We introduce a non-linear Laplacian op...
38. Adaptive Quantum Physics-Informed Neural Networks for Differential Equations with Applications to Fluid Dynamics ​
Author: Fabio Pereira dos Santos, Renato Portugal, J'ulio de Castro Vargas Fernandes, Lucas Timotheo Sanches
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn
arXiv:2608.00850v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have emerged as a versatile approach for solving nonlinear partial differential equations (PDEs), yet achieving high accuracy efficiently using these techniques remains challenging for high-dimensional or multis...
39. HyperODE: Zero-Shot Surrogate for Simulation and Inference of Dynamical Systems ​
Author: Ajitesh Srivastava
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00852v1 Announce Type: new Abstract: Understanding and controlling complex dynamical systems often requires executing thousands of numerical simulations across vast parametric landscapes, which is time-consuming. Machine learning surrogates significantly accelerate simulation by predictin...
40. SparseKAN: Compressing Kolmogorov--Arnold Networks Across Basis Functions, Neurons, and Bits ​
Author: Kazi Ahmed Asif Fuad, Lizhong Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00859v1 Announce Type: new Abstract: Kolmogorov--Arnold Networks (KANs) replace scalar edge weights with learnable univariate functions parameterized by multiple basis coefficients. This introduces a source of redundancy that conventional neural-network compression does not directly expos...
41. Kilobyte Models: Neural Networks as a Seed and a Quantized Latent ​
Author: Sahil Rajesh Dhayalkar
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00860v1 Announce Type: new Abstract: The cost of storing and transmitting a trained neural network scales with its parameter count, a bottleneck for over-the-air updates, on-device libraries, and other bandwidth-bound deployments. We study an extreme form of model compression in which the...
42. GeoArbiter: Verifiability-Guided Grounding for Remote-Sensing Multimodal LLMs ​
Author: Xuechen Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00877v1 Announce Type: new Abstract: Remote-sensing multimodal large language models (MLLMs) often assert facts that imagery cannot establish, such as a facility's identity or function. Coordinate-keyed geographic retrieval can supply this missing knowledge, improving fMoW land-use accura...
43. AOSpec: Action and Observation Co-Speculation for Low-Latency Agent Serving ​
Author: Hao Mark Chen, Jinnan Guo, Wayne Luk, Hongxiang Fan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00881v1 Announce Type: new Abstract: Large language model agents increasingly act through stateful tools, yet model generation and environment execution remain serialized at every step. As decoding accelerates, tool execution becomes a growing bottleneck. Existing action- or observation-o...
44. UpliftBench: Revealing Outcome-Regime and Objective Mismatch in Uplift Evaluation ​
Author: Binshuang Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00915v1 Announce Type: new Abstract: Uplift modeling (conditional-average-treatment-effect estimation) drives personalized targeting, yet published uplift benchmarks frequently disagree on which estimator performs best; we show the disagreement is substantially about metrics, not models. ...
45. Subtype Robustness Is Not Just Accuracy: Calibration Under Unseen Subtype Shift ​
Author: Hanyu Su, Carlota Julbe i Juanola, Yibo Hu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00928v1 Announce Type: new Abstract: Subtype robustness asks whether a model keeps the correct coarse prediction when test examples come from fine-grained subtypes absent from training but still inside a known coarse category. Prior work studies this almost entirely through accuracy. We a...
46. xMICD: Explainable Representation of Multiple ICD Codes ​
Author: Pat Vatiwutipong, Kumkup Keeratisiwakul, Albert Phuoc Kien Van Truong, Nutcha Yodrabum, Wasin Pansiritanachot, Marvin N. Wright, Thanapon Noraset
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00935v1 Announce Type: new Abstract: Electronic Health Records (EHRs) are widely used for clinical risk prediction using machine learning. International Classification of Diseases (ICD) codes provide structured information about patient diagnoses, but representing them effectively remains...
47. Data-Driven Pinball-Loss Selection for Vertically Distributed Elastic-Net SVMs ​
Author: Xiaofei Wu, Kai Qi, Rongmei Liang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.CO
arXiv:2608.00949v1 Announce Type: new Abstract: The pinball-loss support vector machine is robust, but its asymmetry parameter is usually fixed in advance. We propose a data-driven elastic-net support vector machine that learns simplex-constrained weights over candidate pinball losses while retainin...
48. Interpretable machine learning for predicting splitting strength of asphalt concrete: insights from SHAP analysis ​
Author: Jianglei Xing, Xiao Tan, Dongzhao Jin, Pengwei Guo, Yuhuan Wang, Huiya Niu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.00956v1 Announce Type: new Abstract: This paper presents an interpretable machine-learning framework for predicting the splitting strength (ST) of asphalt concrete and supporting data-driven mixture design. A database consisting of 296 samples was established, and 14 input variables relat...
49. One-Sided Quantile Coupling for Flow Matching ​
Author: Jin-Young Kim, So-Yoon Cho, Hyun-Gyoon Kim
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.00978v1 Announce Type: new Abstract: Flow Matching trains continuous-time generative models by regressing the velocity field of a probability path between a simple source distribution and a target data distribution. The coupling that pairs source and target samples strongly affects optimi...
50. Beyond Gene Reconstruction: Learning Cell Representations through Complementary Transcriptomic Views ​
Author: Jiaqi Xiong, Yuntao hu, Yu Zheng, Yifei Shi, Xinyue Guo, Jiaxin Qi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.GN
arXiv:2608.00985v1 Announce Type: new Abstract: The rapid growth of single-cell transcriptomic data has enabled the development of foundation models pretrained primarily by reconstructing masked expression values. This objective encourages these models to learn gene dependencies but does not directl...
51. Who Belongs in the Eval Set? A Capability-Taxonomy-Driven Pipeline for Curating Regression Eval Sets in Agent-Extensibility Platforms ​
Author: Tezan Sahu, Aritra Das, Pankaj Mittal, Sudipta Das
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE
arXiv:2608.01004v1 Announce Type: new Abstract: Platform teams hosting agent-extensibility surfaces face a regression-economics paradox: every onboarding customer ships an evaluation set tuned to their domain, but the platform's regression set must live under a hard query-count ceiling bounded by re...
52. Hierarchical Solomonoff Induction: An Unbounded Machine Learning Model ​
Author: Nathan Young
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT
arXiv:2608.01005v1 Announce Type: new Abstract: Solomonoff Induction, or SolInd, provides an ideal unbounded model of a priori sequence prediction but cannot naturally describe extrapolation from a given training dataset, as performed by Large Language Models. We apply de Finetti's theorem on exchan...
53. Fused Bayesian Flow Networks for Dual-Target Molecular Design ​
Author: Jingyuan Zhou, Shikui Tu, Lei Xu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01007v1 Announce Type: new Abstract: Dual-target drug design aims to generate 3D molecules that can simultaneously interact with two target proteins, offering a promising route for discovering polypharmacological compounds against complex diseases. While recent generative models have show...
54. Caliber: Cross-Architecture Extraction-Cost Control for Score-Returning APIs ​
Author: Chi Wang, Hanwen Wang, Yu Xia, Zihan Wang, Guangdong Bai
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01023v1 Announce Type: new Abstract: We present Caliber, an output-perturbation defense against model extraction that formulates noise selection as a calibration problem: how much the defense degrades the supervision signal used to train a surrogate, and the provable per-input query cost ...
55. The Fourth Quadrant: A Stylized View of Benign Misfitting ​
Author: Gireeja Ranade, Anant Sahai
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2608.01032v1 Announce Type: new Abstract: Training error is what we can observe on a training set; test error is the quantity we actually care about. We study linear regression with squared-error in a deterministic $(d+1)$-dimensional single-spike model. Each stylized training vector has the s...
56. Characterizing Bias in Post-Bandit Inference under Index Algorithms ​
Author: Lisu Wang, Yilun Chen, Jiaqi Lu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, math.PR, stat.ML
arXiv:2608.01069v1 Announce Type: new Abstract: Bandit algorithms generate data for downstream inference, but adaptive sampling biases post-bandit sample means. We analyze this bias for stable index algorithms, including UCB1 and its generalizations, and derive sharp leading-order expressions for th...
57. Logit-Origin Centering for Singleton Test-Time Adaptation ​
Author: Mayank Sharma, Rohit Kumar Mourya, Pratik Mazumder
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.01074v1 Announce Type: new Abstract: Tabular data is used extensively in many real-world use cases. Deep learning models have been developed to deal with tabular data, but generally perform poorly when the test data distribution differs from that of the training data. Researchers have pro...
58. Breaking Diversity Collapse in Spiking Pseudo-Ensembles for Efficient OOD Detection in Remote Sensing ​
Author: Srinivas Anumasa, Rushi Shah, Qiran Zou, Dianbo Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01090v1 Announce Type: new Abstract: Spiking Neural Networks (SNNs) are attractive for resource-constrained remote-sensing systems, but reliable out-of-distribution (OOD) detection remains challenging. Deep ensembles provide strong predictive uncertainty, yet require multiple complete mod...
59. Factorized AdaBoost.MH Achieves the Same Convergence Rate as AdaBoost.MH ​
Author: Xin Zou, Jingyuan Xu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01091v1 Announce Type: new Abstract: AdaBoost.MH reduces multi-class classification to a collection of binary subproblems and enjoys the classical boosting-type convergence guarantee under a weak learning condition. A more structured variant, Factorized AdaBoost.MH, uses base classifiers ...
60. FL-OA: A Byzantine-Robust Federated Learning Framework with Outsourced Auditing for Intelligent Devices ​
Author: Hongliang Zhang, Zhongyuan Yu, Fenghua Xu, Teng Hu, Jian Meng, Jiguo Yu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01095v1 Announce Type: new Abstract: Federated learning (FL) enables multiple intelligent devices to collaboratively train a high-accuracy model without sharing raw data. However, due to its distributed nature, FL is vulnerable to Byzantine attacks. Existing defense methods rely on strong...
61. When Do Surrogate Updates Improve Decisions? A Local Theory of Trajectory-Wise Transfer ​
Author: Yuyang Shen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01130v1 Announce Type: new Abstract: A broad range of models face the mismatch where they are updated through trajectory losses but are evaluated by downstream task reward. Here, a trajectory is a training instance that induces a surrogate loss whose reduction might not track the model's ...
62. Policy Optimality Measurement for Multi-Vehicle Decision-Making: From Extrinsic Indicators to Intrinsic Quality ​
Author: Ye Han, Lijun Zhang, Dejian Meng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01133v1 Announce Type: new Abstract: Evaluating Multi-Agent Reinforcement Learning (MARL) policies in autonomous driving fundamentally relies on extrinsic statistical indicators (e.g., reward curves and success rates), which often mask intrinsic policy degradation and algorithmic blind sp...
63. EulerLoRA: Rank-Driven Jump Dynamics for Calibrated Parameter-Efficient Fine-Tuning ​
Author: Srinivas Anumasa, Dianbo Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01142v2 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning, but standard LoRA produces a single deterministic model and does not directly support predictive uncertainty estimation. We introduce EulerLoRA, a stochastic extension of LoRA that gen...
64. Differentiable Lifting for Topological Neural Networks ​
Author: Jorge Luiz Franco, Gabriel Duarte, Alexander Nikitin, Moacir Ponti, Diego Mesquita, Amauri H. Souza
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2608.01160v1 Announce Type: new Abstract: Topological neural networks (TNNs) enable leveraging high-order structures on graphs (e.g., cycles and cliques) to boost the expressive power of message-passing neural networks. In turn, however, these structures are typically identified a priori throu...
65. Interpretable Machine Learning for Traffic Congestion Prediction: Unveiling the Impact of Different COVID-19 Periods ​
Author: Dan Zhu, Chi Sin Ng, Litian Xie, Yang Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01180v1 Announce Type: new Abstract: Traffic congestion prediction is essential for congestion mitigation, but the COVID-19 pandemic and related control measures altered travel behavior and increased prediction complexity. This study predicts congestion in Alameda County, California, duri...
66. SAFE-Merge: Data-Free Continual Model Merging with General Knowledge Preservation ​
Author: Zihuan Qiu, Zhiyang Liao, Chiyuan He, Yi Xu, Fanman Meng, Linfeng Xu, Qingbo Wu, Hongliang Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01184v1 Announce Type: new Abstract: Data-free continual model merging must incorporate a stream of specialized models while retaining both pretrained general knowledge and previously acquired tasks, without access to task data. Existing methods mainly merge task updates by suppressing in...
67. ReBRAC-v2: The Return of the King ​
Author: Denis Tarasov, Robert K. Katzschmann
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01205v1 Announce Type: new Abstract: Recent offline reinforcement learning methods increasingly rely on expressive generative policies and specialized value-guidance mechanisms. We ask whether comparable progress can instead come from systematically modernizing a conventional behavior-reg...
68. AdaHAT: Adaptive Hard Attention to the Task in Task-Incremental Learning ​
Author: Pengxiang Wang, Hongbo Bo, Jun Hong, Weiru Liu, Kedian Mu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01252v1 Announce Type: new Abstract: Catastrophic forgetting is a major problem in task-incremental learning, where neural networks tend to overwrite previously learned knowledge when trained on new tasks. A number of architecture-based approaches have been proposed to address this proble...
69. Distill What the Student Can See: Fisher-Projected On-Policy Distillation for Vision-Language Models ​
Author: Leyan Xue, Feng Xiong, Mingjun Ma, Changqing Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01263v1 Announce Type: new Abstract: On-policy distillation (OPD) samples trajectories from the current student policy and minimizes token-level divergence between student and teacher next-token distributions at prefixes along those trajectories. This aligns the distillation states with t...
70. Riemannian Attention Mechanisms for Transformers: A Theoretical Framework and Architecture Design ​
Author: Sen Song
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01283v1 Announce Type: new Abstract: All Transformer-based large language models compute attention via the Euclidean inner product, an architectural choice that Dong et al. (2021) proved causes representational rank to decay doubly exponentially with depth in pure self-attention stacks. W...
71. Training nGPT ​
Author: Ilya Loshchilov, Boris Ginsburg
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01284v1 Announce Type: new Abstract: The normalized Transformer (nGPT) realizes hyperspherical representation learning by constraining model parameter vectors and activation vectors to the unit hypersphere. In this paper, we describe a practical training recipe for nGPT and evaluate it on...
72. Stop When Memory Suffices: Evidence-Conditioned Progressive Execution for LLM Agents ​
Author: Yidan Lin, Kaixiang Wang, Jiong Lou, Jie Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01285v1 Announce Type: new Abstract: The continued development of LLMs toward persistent and adaptive intelligence increasingly requires long-term memory mechanisms that preserve and reuse information across interactions. Existing memory systems either compress and structure histories for...
73. FedChronos: Federated Fine-Tuning of Time-Series Foundation Models for Privacy-Preserving Commodity Price Forecasting ​
Author: Amit Sharma, Nitin Auluck, Akramul Azim
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2608.01290v1 Announce Type: new Abstract: Time-series foundation models (TSFMs) such as Chronos have demonstrated strong forecasting capabilities across domains, yet adapting them to institutionally fragmented settings, where data cannot be centralized due to regulatory, competitive, or sovere...
74. AlphaG-OPD: Reliability-Gated Sibling Counterfactuals for On-Policy Distillation in Symbolic Alpha Factor Discovery ​
Author: Yaoyu Su
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01303v1 Announce Type: new Abstract: Symbolic alpha factor discovery can score a completed expression, but it provides no direct label for the structural decisions that produced it. Generative flow networks (GFlowNets) preserve a diverse, reward-proportional distribution over complete exp...
75. Spatiotemporal Proximal Causal Inference under Hidden Confounding and Interference ​
Author: Omar Faruque, Pavan Raj Ravi, Jianwu Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01352v1 Announce Type: new Abstract: Estimating causal effects from real-world spatiotemporal data is challenging due to hidden confounders and interference. Standard causal identification methods assume conditional exchangeability given observed covariates, which fails whenever hidden co...
76. Do Neural Networks Really Beat the Curse of Dimensionality? A Bit-Complexity View ​
Author: Tong Mao, Jinchao Xu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01357v1 Announce Type: new Abstract: Traditional approximation theory measures convergence rates in terms of the number of parameters or degrees of freedom. However, practical computation operates under finite precision: parameters must be encoded using a finite number of bits. Therefore,...
77. When May a Model Replace the Experiment? Audits, Licenses, and the Price of Trust in Surrogate-Driven Design ​
Author: Shuangxiu (Max), Ma (Zachary), Wenhe (Zachary), Zhao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01378v1 Announce Type: new Abstract: Design campaigns in chemistry, materials science, and machine learning share a bottleneck: determining how good a candidate truly is requires an expensive evaluation - an experiment, a first-principles simulation, or a full training run. Machine-learni...
78. On the Identifiability of Masked Prediction: Mode Blindness and Mask Schedules ​
Author: Yichao Cai, Javen Qinfeng Shi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2608.01383v1 Announce Type: new Abstract: Masked prediction learns representations by fitting a schedule-weighted family of conditional laws, but it remains unclear when near-optimal conditional prediction pins down the underlying joint law. We study this question for data with two well-separa...
79. TabDPT-Turbo: Efficient In-Context Learning for Tabular Prediction ​
Author: Rasa Hosseinzadeh, Alex Labach, Zexin Xue, Shuyi Han, Valentin Thomas, Anthony L. Caterini
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01400v1 Announce Type: new Abstract: Tabular foundation models, driven by in-context learning, have rapidly grown in quality and popularity. However, recent approaches with either cell-based architectures or retrieval have sacrificed efficiency for raw performance, restricting their utili...
80. Cluster-Aware Over-the-Air Federated Learning with Energy-Harvesting Devices: From Global Training to Model Personalization ​
Author: Furkan Bagci, Busra Tegin, Mohammad Kazemi, Tolga M. Duman
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2608.01426v1 Announce Type: new Abstract: Federated learning (FL) enables distributed optimization and learning across decentralized edge devices while preserving data privacy, but its performance is fundamentally constrained by heterogeneous data distributions, limited communication resources...
81. Statistical Mechanics of Learning on Product Wasserstein Manifolds ​
Author: Srinivasa Rao P Vangmayi P Reddy
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.NC
arXiv:2608.01434v1 Announce Type: new Abstract: Normally the statistical mechanics of learning treats constraints on weight distributions as restrictions that shrink the space of possible solutions. Therefore, it reduces model capacity. In this paper we would like to take a contrary approach, which,...
82. Conformalized Large Language Models under Configuration Shift ​
Author: Yuqicheng Zhu, Jialin Yu, Lin Li, Gengyuan Zhang, Zhen Yang, Steffen Staab, Puneet Dokania, Philip Torr, Jie Tang, Evgeny Kharlamov
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01460v1 Announce Type: new Abstract: Conformal prediction (CP) is a distribution-free framework for uncertainty quantification that has recently been adapted to large language models (LLMs), providing prediction sets with finite-sample coverage guarantees under exchangeability. Yet for LL...
83. Plasticity of Growing and Elastic Neural Networks in Online Continual Learning ​
Author: Jeong Min Kong, Richard S. Sutton
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01475v1 Announce Type: new Abstract: Neural networks that can grow or both grow and shrink during learning, referred to as growing neural networks and elastic neural networks, respectively, have recently been explored in offline continual learning with a particular focus on catastrophic f...
84. Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval ​
Author: Ilia Semenkov, Daria Kleeva, Ivan Dakhtin, Zarina Maksudova, Alex Ossadtchi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.SD, q-bio.NC
arXiv:2608.01481v1 Announce Type: new Abstract: Short segments of perceived speech can be retrieved from non-invasive magnetoencephalographic (MEG) recordings by deep networks trained with a CLIP-style objective against wav2vec 2.0 audio embeddings. Yet their weights do not map onto electrophysiolog...
85. BiKAN: Restoring Collapsed Basis of Binary Kolmogorov--Arnold Networks ​
Author: Kazi Ahmed Asif Fuad, Lizhong Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01490v1 Announce Type: new Abstract: Binarizing a polynomial Kolmogorov--Arnold Network (KAN) not only changes parameter precision, but also alters the function space available to each layer. When activations are restricted to ${-1,+1}$, all even powers reduce to $1$ and all odd powers re...
86. Stochastic Sequential Search in Very-High-Dimensional Feature Selection ​
Author: Petr Somol, Ji\v{r}'{\i} Grim
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, stat.ML
arXiv:2608.01502v1 Announce Type: new Abstract: Sequential subset search -- forward selection with floating backtracking and its descendants -- remains the quality reference in feature selection, but every member of the family sweeps the full pool of remaining candidate features at each step, which ...
87. Question Begets Question: Self-Evolving Curriculum for Reinforcement Fine-Tuning on Competition Mathematics ​
Author: Longtian Bao, Jianyou Wang, Yang Zhang, Youze Zheng, Ramamohan Paturi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.01522v1 Announce Type: new Abstract: Teaching a language model a skill it has not mastered is obstructed by three recurring difficulties: training data is scarce, ground-truth reasoning traces are usually unavailable, and models often exhibit an apparent ceiling beyond which additional da...
88. Gram-Space: Structure-Preserving Codebook Compression for Memory-Efficient Neuro-Symbolic AI ​
Author: Weilun Wang, Wantong Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01528v1 Announce Type: new Abstract: Vector symbolic architectures (VSA) are widely used for reasoning in neuro-symbolic (NeSy) AI, yet high-dimensional codebooks often create severe memory bottlenecks that limit scalability and deployment. In this paper, we propose Gram-Space, a compress...
89. Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning ​
Author: Seongyoon Kim, Boryeong Cho, Jihwan Oh, Seokhyun Chung, Se-Young Yun
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01556v1 Announce Type: new Abstract: Large language models are increasingly aligned to human preferences via reward modeling, but user preference data are sensitive and often cannot be centralized. Federated learning keeps such data local while learning a shared initial reward model, whic...
90. Meganeura: Portable GPU Training and Inference through Vulkan and Metal ​
Author: Dzmitry Malyshau
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.PL
arXiv:2608.01563v1 Announce Type: new Abstract: Training and deployed inference often cross export, conversion, and platform-specific runtime boundaries. Meganeura asks whether one compact native compiler can span both phases on consumer GPUs. Its typed static graph, automatic differentiation, optim...
91. Measuring in-context algorithmic reasoning in language models against an exact Bayes-optimal standard ​
Author: Hector Zenil, Luan Ozelim
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01575v1 Announce Type: new Abstract: Whether large language models perform genuine algorithmic reasoning or mere pattern completion is hard to test, because most benchmarks lack a ground truth for correct inductive inference. We introduce F-ICL, an in-context-learning benchmark that suppl...
92. HindSearch: Trajectory-Level Hindsight Critique for Search-Augmented Reinforcement Learning ​
Author: Haowei Liu, Jiamian Wang, Hsin-Tai Wu, Zhiqiang Tao, Yi Fang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR
arXiv:2608.01597v1 Announce Type: new Abstract: Search-augmented LM agents are typically trained with a binary exact-match reward, which throws away most of what a failed trajectory tells us about why it failed. We introduce HindSearch, a hindsight self-distillation procedure for GRPO: after each ro...
93. Latent-Regime Bias Auditing for Volatility Forecasting ​
Author: Arthur Chagas, Pedro Bento, Yan Aquino, Arthur Buzelin, Wagner Meira Jr., Cristiano Arbex Valle
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01599v1 Announce Type: new Abstract: Volatility forecasts are commonly evaluated with aggregate accuracy metrics such as RMSE and MAE, but these metrics can hide conditional failures that matter for risk management. This paper proposes a model-agnostic audit framework for evaluating wheth...
94. Online Algorithms via Minimax and Posterior Matching ​
Author: Thomas Kesselheim, Marco Molinaro, Kalen Patton, Sahil Singla
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.DS
arXiv:2608.01616v1 Announce Type: new Abstract: Competitive analysis is central to the study of online algorithms, but upper bounds are often highly problem-specific. We develop a more unifying methodology via the minimax viewpoint. Guided by Yao's principle, we reduce worst-case competitive analysi...
95. QWRF-Net: A Quantum-Wavelet Framework with Rectified Flow for Short-Term Precipitation Nowcasting ​
Author: Zhuo Wang, Chaorong Li, Wenjie Luo, Chuanhu Deng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01626v1 Announce Type: new Abstract: Short-term precipitation nowcasting is important for hydrometeorological early warning, especially when intense convective rainfall may trigger urban flooding, flash floods, and other high-impact hazards. A key challenge in warning-oriented nowcasting ...
96. GraphIR: Architecture-Level Search States for LLM-Guided Neural Architecture Evolution ​
Author: Zhen Liu, Wanqi Zhou, Shuanghao Bai, Yuhan Liu, Jinjun Wang, Jingwen Fu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01633v1 Announce Type: new Abstract: Large language models (LLMs) enable neural architecture search (NAS) directly over executable neural network programs. However, code-level flexibility does not provide the architecture state needed for effective mutation: LLMs must infer tensor depende...
97. Evaluating Forecasting Techniques for Hardware Errors on a Large-scale HPC System ​
Author: Kaiyuan Liao, Xiwei Xuan, Tanwi Mallick, Kevin Brown, Christopher D. Carothers, Kwan-Liu Ma
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01648v1 Announce Type: new Abstract: Hardware error logs in high-performance computing (HPC) systems provide early signals of abnormal behavior, yet there remain challenges in effectively forecasting these errors using modern predictive methods. This work investigates the boundaries of ap...
98. Sharp Root Anti-Concentration via Projective Incidence and Ordered Root Laws ​
Author: Zijun Wang, Yuchen Miao, Yifan Hu, Huanmin Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01670v1 Announce Type: new Abstract: This paper answers the one-dimensional local root anti-concentration questions posed by Balcan, Pegden, and Sharma in the context of online optimization of piecewise-Lipschitz functions. For a homogeneous feature curve and coefficients whose density re...
99. Progressive Agent Skill Generation via Reinforcement Learning ​
Author: Junhao Shen, Zhanqiu Zhang, Yiwen Guo, Hong Cheng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.01678v1 Announce Type: new Abstract: Existing skill generation methods largely rely on heuristics or pipeline-style consolidation, which must be specially designed for different evidence sources. In contrast, learning-based approaches offer a more unified way to model skill generation acr...
100. Beckmann Transport Models: From Autonomous Flows to One-Step Maps ​
Author: Lee Cheuk-Kit, Florentin Coeurdoux, Peter Potaptchik, Yilun Du, Michael Samuel Albergo, Eric Vanden-Eijnden
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01692v2 Announce Type: new Abstract: We propose an instantiation of flow matching that relies on a time-independent velocity field (an \emph{autonomous flow}) to exactly map between two distributions, so long as the target is singular, i.e.\ supported on a lower-dimensional data manifold....
101. Beyond On-Policy Exploration: Integrating External Policy Rollouts for Reinforcement Learning in Diffusion Language Models ​
Author: Wonseok Lee, Jimyeong Kim, Jungmin Ko, Wonjong Rhee
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01717v1 Announce Type: new Abstract: Recent reinforcement learning methods for diffusion large language models (dLLMs) commonly rely on on-policy rollouts generated by the target dLLM itself. When successful on-policy rollouts are scarce, however, on-policy training may receive little pos...
102. LLM-Guided Retrieval for Prediction of Molecular Perturbation Responses ​
Author: Betty Xiong, Jan-Christian Huetter, Gabriele Scalia, Tommaso Biancalani, Sepideh Maleki
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01734v1 Announce Type: new Abstract: Predicting transcriptomic responses to small-molecule perturbations across cell lines is central to drug discovery, but exhaustive profiling of drug-cell combinations is infeasible. We frame molecular perturbation prediction as retrieve-and-aggregate: ...
103. Disagree to Accelerate: Closing the Loop on Diffusion Feature Forecasts ​
Author: Yanchao Li, Jiaqing Xie, Ben Gao, Wanhao Liu, Yanbo Wang, T. Y. Tsui, Jinfei Liu, Yuqiang Li, Tianfan Fu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01740v1 Announce Type: new Abstract: Training-free feature forecasting accelerates diffusion sampling by predicting features at skipped denoising steps. Recent work has mainly focused on designing stronger forecasters. Yet forecast error varies sharply across steps, and open-loop caches t...
104. Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning ​
Author: Li Wang, Xiaodong Lu, Xiaohan Wang, Jiajun Chai, Wei Lin, Tianhao Peng, Guojun Yin
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.01743v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a central paradigm for large language model (LLM) post-training, but optimization toward new objectives can degrade capabilities already present in the base model. KL regularization is widely used to mitigate such...
105. Heterogeneous Multi-Agent Reinforcement Learning for Radio Resource Management under Coupled Finite-Horizon Constraints ​
Author: Yeonseo Jeong, Wonhyeok Ko, Sungweon Hong, Songnam Hong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01745v1 Announce Type: new Abstract: Maximizing throughput under proportional fairness in dense wireless networks requires jointly managing user association, scheduling, base station (BS) activation, and handover control under hard finite-horizon energy and handover budgets, which induces...
106. Multi-Source Dynamic Graph Learning for Compound-Flood Forecasting in Managed Coastal Systems ​
Author: Liangjun You, Min Wu, Orlando Woods, Dongsheng Luo
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01775v1 Announce Type: new Abstract: Compound flooding in managed coastal systems is influenced by hydrological conditions and water-management activity observed across multiple monitoring stations. Current forecasting models can capture temporal dependencies with low average errors, but ...
107. ReFP-AD: Rectified Flow Preconditioning for Energy-Based Anomaly Detection ​
Author: Camile Lendering, Erkut Akdag, Joaqu'in Figueira, Egor Bondarev
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01793v1 Announce Type: new Abstract: Unified anomaly detection requires modeling highly heterogeneous normal data without access to anomalous samples. While foundation models like DINOv2 provide rich token representations, leveraging these spaces for explicit density estimation remains ch...
108. LEAP: Lean Environment-Feedback via Adaptive Pruning for Code RL in GPU Kernel Generation ​
Author: Tankun Li, Zhi Chen, Yaohua Tang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01804v2 Announce Type: new Abstract: Post-training large language models (LLMs) via reinforcement learning (RL) has significantly advanced code generation capabilities. To bypass the heavy memory footprint of critic networks, current state-of-the-art frameworks leverage critic-free paradi...
109. Predictive Maintenance: Deep Learning-Based Remaining Useful Life Prediction for Combat Aircraft Engines ​
Author: Fatih "Urgen, Do\u{g}ay Alt{\i}nel
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY
arXiv:2608.01819v1 Announce Type: new Abstract: To improve the operational readiness of combat aircraft engines and reduce unplanned maintenance costs, accurately estimating the remaining useful life (RUL) is critical. Traditional maintenance often proves insufficient under dynamic mission profiles....
110. tFUSOperator: Operator Learning for Transcranial Focused Ultrasound Digital Twins ​
Author: Minjee Seo, Haris Ghafoor, Minju Seol, Seonaeng Cho, Kyungho Yoon
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.01839v1 Announce Type: new Abstract: Transcranial focused ultrasound (tFUS) requires accurate estimation of the intracranial acoustic field, which is distorted by skull-induced aberrations. Numerical solvers are accurate but computationally expensive for digital twins, where the field mus...
111. WorldDynCache: Risk-Controlled Latent Dynamics Approximation for Diffusion World Model ​
Author: Leyang Chen, Junyi Wu, Shaoqiu Zhang, Yulun Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.01845v1 Announce Type: new Abstract: Diffusion world models generate high-quality futures, but re- peated transformer evaluations make inference prohibitively slow. Existing caches reuse intermediate features, selectively update tokens, or reuse and extrapolate denoising outputs ac- cordi...
112. Beyond Magnitude and Shape: A Direction-Aware Loss for Time Series Forecasting ​
Author: Seunghan Lee, Jaehoon Lee, Jun Seo, Junhyeok Kang, Sangjun Han, Sungdong Yoo, Minjae Kim, Tae Yoon Lim, Dongwan Kang, Hwanil Choi, Soonyoung Lee, Wonbin Ahn
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01857v1 Announce Type: new Abstract: The direction of change --- whether a series will move up or down --- is often as important as its exact value in decisiondriven applications such as risk management and financial forecasting. However, most forecasting losses optimize either point magn...
113. LAB-Tab: LLM-Augmented Bayesian Network Adaptation for Few-Shot Tabular Generation ​
Author: Zijian Shen, Taijie Chen, Bin Zhou, Ziyang Jiang, Jintao Ke
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01879v1 Announce Type: new Abstract: Tabular data generation supports analysis and decision-making when target-domain data are scarce, yet collecting complete target samples is often costly. A practical but underexplored setting provides only a few target records together with richer sour...
114. CARE: A Cascaded Framework for Efficient and Reliable Time Series Anomaly Detection ​
Author: Zemin Chao, Qianhui Xu, Jianhe Cen, Guangzhi Ge, Xiao Chen, Hoangzhi Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01885v1 Announce Type: new Abstract: While deep learning models have achieved state-of-the-art performance in time series anomaly detection, their complex architectures incur substantial inference overhead. Existing methods typically apply a uniform inference strategy across all data poin...
115. Understanding and Correcting Low-Frequency Bias in EEG Foundation Model ​
Author: Junjie Yu, Zihan Deng, Jianyu Zhang, Junrong Mu, Jiahui An, Wenxiao Ma, Ziling Lu, Yue Wang, Yan Zhu, Kexin Lou, Quanying Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01898v1 Announce Type: new Abstract: Increasing EEG pretraining data scale or model capacity does not consistently improve downstream performance. We identify a persistent low-frequency bias in representations learned by diverse EEG foundation models, which remains across dataset scales, ...
116. Finite-Time Analysis of Discounted Exponential-Utility Reinforcement Learning ​
Author: Ankur Naskar, Vivek T A, Aditya Kumar, Gugan Thoppe, Prashanth L. A
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01917v1 Announce Type: new Abstract: Discounted exponential utility provides a principled criterion for risk-sensitive sequential decision-making, but its nonlinear structure complicates reinforcement learning. A recent work \citep{thoppe2026reinforcement} addressed this difficulty by int...
117. HarnessCompass: Guiding Automatic Harness Evolution toward Generalizable and Effective Agent Harnesses ​
Author: Luan Zhang, Ruochen Zhou, Dandan Song, Zhengyu Chen, Yuhang Tian, Jun Yang, Huipeng Ma, Chenhao Li, Guangyuan Feng, Xudong Li, Yizhou Jin, Yan Xu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.01918v1 Announce Type: new Abstract: Harness design plays a critical role in agent performance by shaping how large language models (LLMs) perceive, reason over, and act within executable environments. Recent work has proposed automatic harness evolution, which iteratively improves the ha...
118. ChaosProbe: A Neurochaotic Lens on Frozen Transformer Input-Embedding Spaces ​
Author: Kunal Kumar Pant, Nithin Nagaraj
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2608.01968v1 Announce Type: new Abstract: Transformer models are most often understood through what they do: their benchmark performance, generation quality, or behavior on downstream tasks. Yet frozen transformer input-embedding spaces may also be examined through their responses to a control...
119. AOS: Adaptive Optimizer Switching via Training-State Signals for Faster Convergence and Better Generalization ​
Author: Alok Kumar Pandey, Umang Chaturvedi, Aatish Rana, Gopi Krishna Nedanuri
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01997v1 Announce Type: new Abstract: Single-optimizer training is a poor fit for the distinct phases of deep network optimization: adaptive methods handle noisy early gradients well but overshoot flat minima, while SGD with momentum generalizes better in the late phase but converges slowl...
120. Scikit-fingerprints: Python library for scikit-learn compatible molecular fingerprints and chemoinformatics ​
Author: Jakub Adamczyk, Adam Staniszewski
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.SE
arXiv:2608.02027v1 Announce Type: new Abstract: We present scikit-fingerprints, a comprehensive, fully scikit-learn compatible library for molecular machine learning in Python, based on RDKit. Molecular fingerprints and related functionalities are workhorses of chemoinformatics, yet the widely used ...
121. DART: Decoded Attention over Recurrent States for Efficient Long-Context Sequence Modeling ​
Author: Yixiao Qian, Song Chen, Pengkai Wang, Jiaxu Liu, Shengze Cai, Chao Xu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02032v1 Announce Type: new Abstract: Modern language models are built primarily from Transformers, recurrent models, and their hybrid architectures. Transformers rely on token-level attention memories, while recurrent models such as state space models (SSMs) and linear attention maintain ...
122. Upper-Expectile Multi-Step Q-Learning for Off-Policy Reinforcement Learning ​
Author: Abdelghani Ghanem, Mounir Ghogho
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02034v1 Announce Type: new Abstract: Multi-step returns accelerate reward propagation in off-policy reinforcement learning, but couple the evaluation of each decision to the suboptimal logged actions that follow it, inducing a pessimistic bias that grows with the horizon. We propose Expec...
123. Convex Neural Energy Elements: Monolithic Finite-Element Assembly of Geometry-Parameterized Neural Operators with Stability and Error Guarantees ​
Author: Hongyue Jiang, Jianjiang Zhan, Chenzhuo Zhang, Fan Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02036v1 Announce Type: new Abstract: Extending the neural-operator element method from individually trained, fixed-geometry neural elements to a library of reusable, geometry-parameterized element types fails structurally: a field-predicting operator trained by value regression induces an...
124. Secrets Everywhere: Auditing Memorization in Mobility Prediction Models ​
Author: Anne Josiane Kouam, Hristo Boyadzhiev, Konrad Rieck
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02052v1 Announce Type: new Abstract: Human mobility prediction models, which forecast the next location in a user's trajectory, are increasingly deployed in urban analytics, navigation, and personalized services. Yet, little is known about their potential to memorize and expose sensitive ...
125. SCOPE: Entanglement Frontier Escape for Source-Free Class Unlearning ​
Author: Junhao Cai, Dohun Kim, Sung Il Choi, Juhyun Park, Chengjun Jin, Dowon Kim, Changhee Joo
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02058v1 Announce Type: new Abstract: Source-free class unlearning erases whole classes using only the forget data, judged at the representation level, where features can leak a class the head no longer predicts. Existing feature-space erasers answer with one fixed projection, yet forget a...
126. Geometry-Guided Layerwise FFN Width Allocation in Transformers ​
Author: Timur Mudarisov, Mikhail Burtsev, Radu State
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.02064v1 Announce Type: new Abstract: Feed-forward networks (FFNs) account for a large fraction of Transformer parameters, yet their hidden width is usually constant across depth. We ask whether this capacity can instead be allocated from a forward-pass measurement of layer behavior. We vi...
127. Feed-Forward Steering in Transformer Residual Dynamics ​
Author: Timur Mudarisov, Mikhail Burtsev, Radu State
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, math.DS
arXiv:2608.02071v1 Announce Type: new Abstract: Attention-only dynamical theories model Transformer residual directions as particles aggregating on a sphere. We extend this framework by incorporating the feed-forward network (FFN) term as a local steering field acting on each token state. The result...
128. Isotonic Bradley-Terry Model for Paired Comparison Data ​
Author: Ryoya Yamasaki
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02081v1 Announce Type: new Abstract: In this paper, we study prediction problems for paired comparison data, for example, predicting the win probability between two unmatched players and ranking all the players according to the order of their strengths by using win probability data betwee...
129. A 2-Block Architecture for Real-Time EEG Gait Decoding: A Pilot Study ​
Author: Shantanu Sarkar, Saurabh Prasad, Jose L. Contreras-Vidal
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, eess.SP
arXiv:2608.02083v1 Announce Type: new Abstract: Closed-loop lower-limb exoskeleton control via Electroencephalography (EEG) remains limited by motion artifacts, low signal-to-noise ratio, and binary gait formulations that fail to capture full cortical gait complexity. We propose a 2-block Brain-Comp...
130. An AI-Based Decision-Support Pipeline for Day-Ahead Photovoltaic Forecasting ​
Author: Fariba Dehghan, Sebastian Stein, Vahid Yazdanpanah, Stephanie Gauthier, Masood Nazari
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY
arXiv:2608.02088v1 Announce Type: new Abstract: Reliable photovoltaic (PV) forecasts are needed for low-carbon energy systems, but newly deployed sites often have short, imperfect records. This makes standard day-ahead forecasting difficult: persistence and physical baselines can be sensitive to cal...
131. How Much Does a Reasoning Summary Reveal? An Observability Ladder for Large Language Models ​
Author: Andres Algaba, Francesca Carlon, Lynn Delcon, Marthe Ballon, Bert Verbruggen, Vincent Ginis
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.02089v1 Announce Type: new Abstract: Large language models often show users a final response and a short reasoning summary while the full reasoning trace stays hidden. We introduce an observability ladder that holds each completed run fixed and varies only what a reader inspects to judge ...
132. One QK Channel, Many Sources: Guarding Low-Precision Attention Collapse ​
Author: Shuxiao Xie, Shuyang Xie, Yuan Cao, Dezhi Ran, Wei Yang, Tao Xie
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02091v1 Announce Type: new Abstract: A bfloat16 transformer can train normally for many steps and then collapse abruptly. Distinct low-precision errors can trigger the same failure, leaving unclear whether each source needs its own repair or one shared route can be blocked. We isolate a r...
133. Do Static Embeddings Add Value to Hybrid Dutch Retrieval? ​
Author: Ant'onio Pereira Barata
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2608.02112v1 Announce Type: new Abstract: Embedding benchmarks measure standalone model quality, but they do not establish whether a low-cost retriever contributes complementary ranking information once lexical and transformer-based retrieval are already combined. We present a controlled evalu...
134. CoRe-GNN: Multilevel Message passing on Coarsened graphs ​
Author: Antonin Joly, Nicolas Keriven, Aline Roumy
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02128v1 Announce Type: new Abstract: Training Graph Neural Networks on large graphs is challenged by the memory cost of storing all node representations across layers. We show that several existing scalable approaches can be written as structured modifications of the GNN propagation matri...
135. RamanPFN: learning from Raman spectral structure with a tabular foundation model ​
Author: Xingyu Pan, Huan Wang, Jinjia Guo, Zhenlin Zhao, Siming Dong, Jixi Lu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.02157v1 Announce Type: new Abstract: Raman spectroscopy enables non-destructive, label-free molecular characterization across materials science, biomedicine and process monitoring. Predictive Raman datasets often contain few labelled spectra and thousands of ordered wavenumbers, with info...
136. Empowering Credit Risk Detection in Weixin Pay with Billion-Scale Deep Graph Learning ​
Author: Xin Liu, Xiyuan Chen, Chenglong Wu, Xuan Zong, Jun Zhou, Dawei Cheng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02168v1 Announce Type: new Abstract: Credit risk detection, particularly mitigating individual fraud, is crucial for maintaining the stability of digital financial ecosystems. Accurately identifying credit fraud among billions of users is critical for minimizing financial losses and safeg...
137. Start Classifying: Categorical Critics for LLM Reinforcement Learning ​
Author: Zhijian Zhou, Long Li, Xuan Zhang, Zongkai Liu, Yulei Qin, Ke Li, Xing Sun, Xiaoyu Tan, Chao Qu, Yuan Qi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02181v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) for large language models typically trains its critic by mean-squared-error (MSE) regression on scalar value targets. Although scalar MSE is statistically valid for estimating the conditional expected return, sparse b...
138. CRIP: Channel Level Representation Injection for Personalized One-Shot Federated Learning ​
Author: Zijian Jiang, Chaoli Sun, Handing Wang, Xilu Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02222v1 Announce Type: new Abstract: One-shot federated learning (OSFL) has emerged as a promising collaborative model learning framework with only a single round of communication, offering significant advantages in communication efficiency and privacy preservation. However, OSFL often fa...
139. Constrained Co-Design for Photonic Bayesian Neural Networks ​
Author: Hendrik Borras, Xiao Wang, Bernhard Klein, Robin Janssen, Frank Br"uckerhoff-Pl"uckelmann, Wolfram Pernice, Holger Fr"oning
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02229v1 Announce Type: new Abstract: Classical neural networks frequently produce overconfident predictions on ambiguous or out-of-distribution (OOD) data, a liability that grows with each AI system deployed in safety-critical real-world scenarios. Bayesian neural networks (BNNs) provide ...
140. Assessing the Impacts of Imperfect Datasets on Client Selections in Federated Learning ​
Author: Yuan-Heng Tsai, Li-Hsing Yen, Yan-Wei Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.02250v1 Announce Type: new Abstract: Federated learning (FL) is a popular distributed learning framework where multiple clients perform local training and a server aggregates the locally updated models. FL enables decentralized training while preserving the privacy of clients' datasets. H...
141. Z-PEFT: Zero-shot Backdoor Detection in Parameter-Efficient Fine-Tuning via Canonical Spectral Signatures ​
Author: Nicola Pitzalis, Donald Shenaj, Giacomo Cignoni, Andrea Cossu, Davide Bacciu, Antonio Carta
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02271v1 Announce Type: new Abstract: Parameter-Efficient Fine-tuned (PEFT) models are frequently downloaded from open repositories by practitioners. This widespread practice creates a significant attack surface, as malicious actors can publish backdoored models that induce specific behavi...
142. BRiG-AFA: Bellman Risk-to-Go Learning for Non-Myopic Active Feature Acquisition ​
Author: Jiaorong Feng, Qian Li, Ying Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.02305v1 Announce Type: new Abstract: Active feature acquisition (AFA) asks which unobserved feature to measure next for each test instance under a budget. Greedy rules are easy to train but can overlook context features whose value is realized only through later acquisitions, while reinfo...
143. Diffusion Policy with Behavioral Advantage Correction for Offline Reinforcement Learning ​
Author: Botao Dong, Longyang Huang, Ning Pang, Hongtian Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY
arXiv:2608.02332v1 Announce Type: new Abstract: In offline reinforcement learning (RL), the distribution shift between behavioral data and the learned policy can lead to erroneous \emph{Q}-value estimation, thereby misguiding the direction of policy optimization. To address this issue, we develop a ...
144. Qwen-CUA: Native Computer Use for (almost) Everything ​
Author: Dunjie Lu, Shuai Bai, Tianyi Bai, Sicheng Fan, Chang Gao, Jian Guan, Feng Hu, Mianqiu Huang, Xingyang Huang, Yizhen Jiang, Yuheng Jing, Dehui Kong, Ning Li, Dayiheng Liu, Shixuan Liu, Zheng Liu, Que Shen, Bowen Wang, Junli Wang, Chencan Wu, Rui Xie, Tianbao Xie, Zhihui Xie, Haiyang Xu, An Yang, Tao Yu, Wenzhen Yuan, Xi Zhang, Zhenru Zhang, Mingkang Zhu, Zhaoqing Zhu, Yizhong Cao, Kai Dang, Binyuan Hui, Kaixin Li, Junyang Lin, Haiquan Wang, Zekun Wang, Yiheng Xu, Fan Yan, Mengqi Yuan, Danyang Zhang, Jiajun Zhang, Zhipeng Zhang, Fan Zhou, Fan Zhou
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.02352v1 Announce Type: new Abstract: Native computer use offers a general interface for agents to operate almost any software available to people, but requires long-horizon state tracking, large-scale interactive experience, and learning from sparse yet verifiable outcomes. We introduce Q...
145. GLAIM: Learning Global and Local Adaptive Inter-Variable Dependency for Multivariate Time Series Imputation ​
Author: Mingyang Wang, Rongwen Li, Xiao Wang, Changjian Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.02366v1 Announce Type: new Abstract: Multivariate time series imputation is fundamental to downstream analysis, yet modeling inter-variable dependencies with incomplete observations remains challenging. Existing methods learn global dependencies across samples or dynamic local dependencie...
146. Gecko: Fast Private Inference via Secure Public Encoder Offloading ​
Author: Cheng'an Wei, Kai Chen, Yue Zhao, Congyi Li, Shenchen Zhu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02378v1 Announce Type: new Abstract: Private inference protects both user inputs and server models during neural network inference, but existing solutions remain too slow for practical deployment. This motivates recent efforts to run a public encoder, such as a pretrained backbone, outsid...
147. From fragmented data to actionable design: Physics-calibrated learning for plastic upcycling ​
Author: Jingyang Bai, Zijia Wang, Xiangyi Long, Marcos Millan, Binjian Nie, Mingyue Ding
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.02402v1 Announce Type: new Abstract: Thermochemical upgrading of plastic waste is a key upcycling pathway, yet the experimental literature is fragmented by heterogeneous conditions and incomplete reporting. Complete-case learning would retain only 10.99% of the curated experiments, while ...
148. Deep Learning-Based Estimation of Ground Reaction Forces in Parkinsonian Gait Using an Optimized Set of IMU Data ​
Author: Run Lin, Yingtian Tang, Jiawen Xu, Dongfei Huo, Lefan Wang, Helen Dawes, Dominic J. Farris, Dong Wang, Xijin Hua
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2608.02408v1 Announce Type: new Abstract: Accurate gait analysis in Parkinson's disease (PD) typically relies on laboratory-based systems to capture biomechanical data, such as ground reaction forces (GRFs). Estimating GRFs using inertial measurement units (IMUs) provides a feasible alternativ...
149. Why Large Language Models Fail at Tabular Prediction ​
Author: Marta Garnelo, Wojciech M. Czarnecki
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02412v1 Announce Type: new Abstract: Large language models (LLMs) have become the default tool for a remarkable range of tasks, yet they have had conspicuously little success at one of the most common machine learning workloads: predictive analytics over tabular data. This gap is the foun...
150. Foundations of Reinforcement Learning and Control:Connections and New Perspectives ​
Author: Claire Vernade, Onno Eberhard, Martha White, Florian D"orfler, Csaba Szepesv'ari, Miroslav Krstic, Michael Muehlebach
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02433v1 Announce Type: new Abstract: Reinforcement learning and control theory are two adjacent scientific fields that focus on optimizing the controller of unknown dynamical systems using feedback. While both fields have common roots in dynamic programming, they have evolved with distinc...
151. Aggregate-then-Calibrate for Human-centered Assessment with Theoretical Guarantees ​
Author: Zejun Xie, Xintong Li, Guang Wang, Desheng Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.02455v1 Announce Type: new Abstract: Human-centered assessment tasks, which are essential for systematic decision-making, rely heavily on human judgment and typically lack verifiable ground truth. Existing approaches face a dilemma: methods using only human judgments suffer from heterogen...
152. RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States ​
Author: Yi Yang, Zhennan Chen, Yihong Zhuang, Tiehan Fan, Yinan Chen, Jian Li, Jian Yang, Ying Tai
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.02508v2 Announce Type: new Abstract: Learning-based memory systems for self-evolving LLM agents face two tightly coupled challenges. First, trajectory-indexed utilities grow with the interaction history, thereby dispersing limited feedback over an ever-expanding state space. Second, becau...
153. Analytic Planning under Uncertainty with Moment Closure ​
Author: Shishir Sharma, Doina Precup
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.02519v1 Announce Type: new Abstract: Effective model-based reinforcement learning in stochastic environments requires planning that accounts for predictive uncertainty. Propagating full state distributions analytically offers a principled way to do this, but has traditionally required res...
154. Uncertainty Is Not Enough: Value-of-Information Routing for Mixtures of LoRA Experts ​
Author: Tom Saliencro, Rohan Desai, Priya Nair, Maya Lindqvist, Daniel Whitmore
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02528v1 Announce Type: new Abstract: Mixtures of low-rank adaptation experts increase parameter-efficient capacity by routing each input through a subset of adapters. Recent dynamic routers activate more experts when the router or prediction is uncertain. This rule silently equates uncert...
155. Benchmarking Sheaf Neural Networks for Inductive Tasks ​
Author: Stefano Fiorini, Edoardo Coppola, Pietro Li`o
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02558v1 Announce Type: new Abstract: Sheaf Neural Networks (SNNs) generalize message passing by replacing scalar edge weights of standard Graph Neural Networks (GNNs) with learnable, edge-dependent restriction maps between node stalks. Despite their strong theoretical foundations and prom...
156. Structured Memory for Edge Language Models: Persistent Context and Corpus Retrieval via O(1) SSM State Injection ​
Author: Anusha Madan Gopal, Aras Pirbadian, Kristofor D. Carlson, M Anthony Lewis, Jonathan Tapson
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR
arXiv:2608.02560v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) imposes a prefill cost proportional to retrieved context length, and -- with Transformer backbones -- a KV-cache that grows with each generated token. State-Space Models (SSMs) avoid the second cost by construction;...
157. Pseudorandom Streams within Diffusion Models Act as Learnable Inputs That Affect Generation Quality ​
Author: Shengzhi Deng, Chenqi Ye, Yanze Guo
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.02575v1 Announce Type: new Abstract: Diffusion models rely on stochastic inputs, yet on finite-precision hardware, the "randomness" they consume is realized as deterministic numerical orbits generated by pseudorandom rules. Accessible orbit structure can become a learnable input and affec...
158. Smooth Reparameterizations of Functions on Simplicial Product Spaces: Applications to Probabilistic Tensor Decomposition and Functional Data Registration ​
Author: Shashwat Kumar, Arafat Rahman, Anuj Srivastava, P. -A. Absil
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02576v1 Announce Type: new Abstract: We consider optimization problems defined on product spaces of simplices. Examples of this class of problems include learning low-rank discrete multivariate probability distributions via simplex constrained tensor decomposition and performing functiona...
159. GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning ​
Author: Zhaoxin Yu, Qi Shen, Hengli Li, Zhaowei Zhang, Song-Chun Zhu, Chi Zhang, Zilong Zheng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.02585v1 Announce Type: new Abstract: Optimization-based latent reasoning improves large language model outputs by optimizing instance-specific continuous states at test time while keeping model parameters frozen. Existing methods, however, typically connect these states to the reasoning t...
160. onepot-Bench 0: towards lab-aware in silico chemistry benchmarks ​
Author: Brandon Wang, Andrei S. Tyrin, Daniil A. Boiko
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02595v1 Announce Type: new Abstract: Language models are playing an increasingly important role in laboratory science, performing tasks such as experiment planning, execution, and post-hoc analysis. However, precisely measuring their abilities is difficult, as scientific capabilities requ...
161. Enriched text-guided variational multimodal knowledge distillation network (VMD) for automated diagnosis of plaque vulnerability in 3D carotid artery MRI ​
Author: Bo Cao, Fan Yu, Mengmeng Feng, SenHao Zhang, Xin Meng, Yue Zhang, Zhen Qian, Jie Lu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2509.11924v2 Announce Type: cross Abstract: Multimodal learning has attracted much attention in recent years due to its ability to effectively utilize data features from a variety of different modalities. Diagnosing the vulnerability of atherosclerotic plaques directly from carotid 3D MRI imag...
162. Amplitude-Only FFN Intervention for Tool-Structured LLM Inference Method: Gated Evaluation Protocol, and Cross-Model Empirical Results ​
Author: Sheng Xu, Junhua Wang, Boyuan Huang, Ke Jia, Jiadun Zhu, Zhen Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.11183v2 Announce Type: cross Abstract: Large language models increasingly operate as tool-using agents, where small format, argument, or function-call errors can invalidate otherwise plausible responses. We study inference-time feed-forward network (FFN) intervention for improving structu...
163. Posterior Variance Is a Constraint Map, Not an Error Map: Closed-Form Uncertainty for Radiative Gaussian Splatting in Sparse-View CT ​
Author: Chulin Zhao, Yiran Xu, Shu Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV
arXiv:2607.13682v2 Announce Type: cross Abstract: Radiative Gaussian splatting reconstructs sparse-view CT fast and accurately, and recent work attaches per-Gaussian posteriors to yield per-voxel uncertainty maps. We ask what such a map actually measures: posterior variance is a data-constraint map,...
164. Learning to Persuade Privately Informed Receivers ​
Author: I. Arda Vurankaya, Ufuk Topcu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2607.28342v1 Announce Type: cross Abstract: Bayesian persuasion studies how an informed sender can influence the behavior of a receiver through strategic information disclosure. Standard models assume the sender is the receiver's only source of information, yet in many applications receivers a...
165. Cost-Effective Automated Judging of Natural-Language Mathematical Proofs ​
Author: Benjamin Grayzel
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.00004v1 Announce Type: cross Abstract: Grading natural-language mathematical proofs is a recurring cost in evaluating math-reasoning systems, and frontier LLM judges are expensive. We ask whether cheap open-weight models can serve as reliable judges given a candidate proof, a ground-truth...
166. MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents ​
Author: Bohan Tang, Yiwen Guo
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.00007v1 Announce Type: cross Abstract: Equipping Large Language Models (LLMs) with human-like personas is crucial for agentic applications, such as role-play and user simulation. Traditional prompt-based methods rely on descriptive conditioning by injecting static textual profiles, which ...
167. Obshazard-bench: Benchmarking Multimodal Foundation Models for Real-Time Disaster Intelligence from Raw Earth Observation Streams ​
Author: Fengxiang Wang, Qiuyang Yu, Yueying Li, Mingshuo Chen, Chengchi Fei, Kaiyi Xu, Lixin Gu, Wangxu Wei, Junchao Gong, Lipeng Ma, Jiong Wang, Fenghua Ling, Wenlong Zhang, Xue Yang, Wenjing Yang, Ben Fei, Long Lan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG
arXiv:2608.00012v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are increasingly used to interpret Earth observation data, yet their capability to support real-world disaster emergency response remains insufficiently evaluated. Existing remote sensing benchmarks largely re...
168. Nova: An End-to-End MLIR Compiler for Deep Learning ​
Author: Adwaid Suresh, Aparna A, Harshini V M, Jona Delcy C A, Killi Uma Maheswara Rao, Ram Charan Golla, Surendra Vendra
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.AR, cs.LG, cs.PL
arXiv:2608.00029v1 Announce Type: cross Abstract: The performance of deep learning models at scale relies heavily on how effectively high-level mathematical operations are mapped to underlying physical hardware. While high-level tensor frameworks provide flexible abstractions for model design, their...
169. Retrieval-Based Cross-Domain Generalization in Optical Networks via Global Features ​
Author: Ali Al Housseini, Carlos Natalino, Paolo Monti, Omran Ayoub
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.CV, cs.IR, cs.LG, cs.NI
arXiv:2608.00044v1 Announce Type: cross Abstract: We propose a retrieval-based framework for crossdomain quality-of-transmission (QoT) estimation that leverages transferable feature representations while avoiding reliance on source-domain-specific decision boundaries. The proposed approach supports ...
170. Not All EEG Moments Are Equal: Position-Adaptive Time Scheduling for EEG Generation ​
Author: Boheng Liu, Ziyu Li, Chenghua Duan, Qing Li, Xia Wu
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG
arXiv:2608.00048v1 Announce Type: cross Abstract: Electroencephalography (EEG) generation is essential for alleviating data scarcity and enabling large scale neural modeling in brain computer interface applications. However, existing flow based approaches assume that every channel and every time seg...
171. Identifiability-Aware Source Apportionment in City-Scale Advection-Diffusion Systems ​
Author: Ankit Bhardwaj, Lakshminarayanan Subramanian
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.CE, cs.LG
arXiv:2608.00050v1 Announce Type: cross Abstract: Source apportionment from sparse urban air-quality sensors is an inverse problem limited by sensor placement, wind-driven transport, background variation, and noise. Known or proxy emission inventories make attribution meaningful by restricting the u...
172. Hybrid-Field Sparse Channel Representation and Recovery for XL-RIS-Assisted mmWave MIMO Systems ​
Author: Wenkai Liu, Nan Ma, Jianqiao Chen, Hongtao Zhang, Ping Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.00052v1 Announce Type: cross Abstract: Extremely large-scale reconfigurable intelligent surface (XL-RIS)-assisted communication is regarded as a key enabling technology for future 6G networks. However, hybrid-field channel estimation for XL-RIS-assisted systems is challenging due to the h...
173. Fast Trainable Multilinear Bases for Image Compression ​
Author: Shiwen An, Zhongyi Ni, Huanhai Zhou, Jin-Guo Liu
Published: 8/4/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, math.OC, quant-ph
arXiv:2608.00053v1 Announce Type: cross Abstract: The Discrete Fourier Transform, the Discrete Cosine Transform, and their block-wise variants underpin most deployed image and video codecs. Their effectiveness rests on three properties: they run in near-linear time (linear up to a polylogarithmic fa...
174. Domain-Generalized Adaptive Semantic Communication for Collaborative Perception ​
Author: Fan Gao, Youzheng Wang, Ning Ge
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.LG, eess.IV
arXiv:2608.00056v1 Announce Type: cross Abstract: We propose RSTA, a domain-generalized semantic communication framework enabling source-free V2X collaborative perception under both observation-domain shift and unseen wireless channel conditions. In V2X, received semantic tokens suffer coupled degra...
175. Automated ECG Interval Measurement and Wave Delineation Using Fast Fourier Convolution ResNet ​
Author: Farhan Adam Mukadam, Harshit Mishra, Nachiket Makwana, Pradyot Tiwari, Subramani Kandasamy, KVS Hari
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG
arXiv:2608.00058v2 Announce Type: cross Abstract: Accurate measurement of ECG intervals, including PR, QRS duration, and QT/QTc, is central to cardiac diagnosis, yet the published ECG delineation literature evaluates performance almost exclusively as fiducial-point timing errors on small curated dat...
176. A Spatial Persistence Gradient in European Warming Consistent with North Atlantic Cold-Blob Influence ​
Author: Mauricio Herrera-Mar'in, Alex Godoy-Fa'undez, Diego Rivera
Published: 8/4/2026, 4:00:00 AM
Categories: physics.geo-ph, cs.LG, physics.ao-ph
arXiv:2608.00063v1 Announce Type: cross Abstract: Europe is warming faster than the global mean, yet the spatial organisation of this acceleration remains incompletely understood. Using ERA5 reanalysis for 1950--2024 across 28 IPCC AR6 European sub-regions, we identify two connected empirical result...
177. H+ Embedding: Harmonizing Global and Token-Level Retrieval with Context-Dependent Phrases ​
Author: Shusen Zhang, Junyi Hu, Ye Feng, Ziteng Wang, Zhaoyuan Pan, Guosheng Dong, Xiaojun Yuan, Jiangshou Hong, Xiangzhi Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.00065v1 Announce Type: cross Abstract: Terminology-intensive retrieval, especially in medical settings, depends on preserving multi-word entities, abbreviations, numerical constraints, and compositional concepts. However, existing representations lie at two extremes: single-vector retriev...
178. Hybrid Quantum CNN for Cross-Sensor Spaceborne Volcanic Thermal Activity Recognition Worldwide ​
Author: Claudia Corradino, Federica Torrisi, Alessandro Grilli, Tommaso Catuogno, Mattia Verducci, Elisabetta Paladino, Luigi Giannelli, Alessandro Sebastianelli
Published: 8/4/2026, 4:00:00 AM
Categories: physics.geo-ph, cs.LG
arXiv:2608.00069v1 Announce Type: cross Abstract: As Earth Observation (EO) enters the Big Data era, the exponential volume of daily satellite imagery poses significant computational and storage challenges for classical Deep Learning (DL) models. Moreover, current approaches often struggle to genera...
179. Beyond Random Partitioning: Unsupervised Spatio-Temporal Stratification for Cohort Balancing in Longitudinal Medical Imaging ​
Author: Qinghui Liu, Jon Andr'e Ottesen, Atle Bj{\o}rnerud, Kyrre Eeg Emblem
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.00073v1 Announce Type: cross Abstract: Rigorous dataset partitioning is a foundational, yet frequently overlooked, prerequisite for reliable deep learning in longitudinal medical imaging. Naively shuffling small clinical cohorts routinely introduces covariate shifts and temporal sampling ...
180. DODA: A Database of Datasets for Aesthetics Research ​
Author: Lisa Ko{\ss}mann, Ralf Bartho, Christoph Redies, Johan Wagemans
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.00089v1 Announce Type: cross Abstract: With rapid growth in the fields of empirical and computational aesthetics we have seen a vast increase in large image datasets annotated for aesthetics. As the image databases differ widely in many respects (e.g., different standards for annotation),...
181. Rethinking Total Absorption Gamma Spectroscopy Deconvolution: Supervised Machine Learning vs Response-Matrix Methods ​
Author: J. Balibrea-Correa, E. N{'a}cher, C. Fonseca-Vargas, J. L. Tain
Published: 8/4/2026, 4:00:00 AM
Categories: physics.data-an, cs.LG, nucl-ex
arXiv:2608.00090v1 Announce Type: cross Abstract: The extraction of $\beta$-feeding distributions in Total Absorption $\gamma$-ray Spectroscopy constitutes a challenging inverse problem, particularly in nuclei with complex decay schemes involving a large number of excited states. In such cases, the ...
182. Conservation laws determine what physical learning remembers ​
Author: Bijaya Dangol
Published: 8/4/2026, 4:00:00 AM
Categories: cond-mat.soft, cond-mat.dis-nn, cs.AI, cs.LG
arXiv:2608.00097v1 Announce Type: cross Abstract: Physical learning rules such as equilibrium propagation (EP), coupled learning (CL), and adjoint coupled learning (AL) train resistive networks through local measurements. In the small-nudge limit EP and CL exactly conserve the conductance mass K = (...
183. Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale ​
Author: Banruo Liu, Haoran Qiu, 'I~nigo Goiri, Rodrigo Fonseca, Ricardo Bianchini, Esha Choukse
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.00101v1 Announce Type: cross Abstract: AI coding agents like GitHub Copilot, Claude Code, and Codex interleave multi-step LLM inference with tool execution, creating a workload different from chatbots. We present the first production-scale characterization of this workload using sampled G...
184. FDIR: Harmonizing Fidelity and Human-Machine Preference in Lossy Compression Image Restoration ​
Author: Kuan-Yen Chen, Fang-Yi Su, Philip Chikontwe, Jung-Hsien Chiang
Published: 8/4/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2608.00111v1 Announce Type: cross Abstract: Image restoration quality can be evaluated along three complementary facets: pixel-level fidelity, human perception, and downstream machine preference. However, existing lossy compression restoration methods optimize for at most one of these criteria...
185. Deep Learning for Cyber Threat Detection and Mitigation in Healthcare-IoT ​
Author: Mirza Akhi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, eess.SP
arXiv:2608.00118v1 Announce Type: cross Abstract: Cybersecurity is a fundamental requirement for protecting wearable devices used in healthcare Internet of Things (H-IoT) systems. Security failures in these resource-constrained systems directly compromise patient safety. Physiological data and netwo...
186. LLM-OSDA: An Optimal-Stopping Dynamic Auction for Native Advertising in Multi-Turn LLM Conversations ​
Author: Yan Fang, Jialin Chen, Chun Gan, Hang Yu, Mingjun Nie, Yeyu Zhang, Fengxiang He, Ching Law
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.GT, cs.LG
arXiv:2608.00123v2 Announce Type: cross Abstract: LLM-native advertising embeds sponsored content directly into model-generated responses, shifting the unit of sale from a fixed slot to a moment within an evolving conversation. Existing LLM ad-auction mechanisms primarily operate within a single res...
187. RadPRISM: Schema-stratified radiology-report supervision for concept-disentangled image representations and visual grounding ​
Author: Fabian Drexel, Marlene Fritzsche, Era Stambollxhiu, Miriam Kumpf, Lena Schmitzer, Lea Schumann, Jannik Kahmann, Friedrich Puttkammer, Johannes Moll, Jannik L"ubberstedt, Zeineb Ben Chaaben, Anirudh Narayanan, Cosmin I. Bercea, Sebastian Ziegelmayer, Marcus R. Makowski, Daniel Rueckert, Lisa C. Adams, Keno K. Bressem
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.00147v1 Announce Type: cross Abstract: Vision-language pretraining learns rich medical image representations from radiology reports, but previous model variants commonly operate within a single shared embedding space, so concept-level structure and interpretability must be recovered post ...
188. Generative Models for Modeling and Synthesizing MIMO Channels in Adverse Weather Conditions ​
Author: Vignesh Nandakumar, Faraz Barati, Brian L. Evans
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.00156v1 Announce Type: cross Abstract: The push for broader coverage in future cellular networks depends on reliable service, yet this is increasingly harder to do as we encounter more instances of extreme weather conditions. In extreme weather conditions, we have difficulty evaluating co...
189. A Constitution-Grid Instrument for Data-Efficient RL Alignment (C-Guard) ​
Author: Xianling Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.00180v2 Announce Type: cross Abstract: Conflicting objectives are general in RL alignment, and training on them data-efficiently is hard. Training a safety guard with RL means optimizing two objectives that conflict: catch real harm, and do not refuse benign prompts. Our finding is that o...
190. SCALP: Semi-Supervised Statistical Shape Modeling from Imperfect 3D Photogrammetry via Landmark-Anchored Spectral Warp ​
Author: Nawazish Khan, Sanjay Bhandari, Sarang Joshi, Alzbeta Novotna, Tiffany Jeong, Loretta Bowman, Michael Hernandez, Tobi Somorin, Viraj Govani, Jesse Glodstein, Shireen Elhabian
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.00187v1 Announce Type: cross Abstract: Correspondence-based statistical shape modeling (SSM) is vital for population-level morphometric analysis, but conventional pipelines assume clean, fully registered surfaces. Real-world clinical photogrammetry scans are often noisy, partial, and clut...
191. MedSAM2-Anatomy: Training-Free Inference-Time Optimization for Musculoskeletal Segmentation ​
Author: John Garcia Henao, Nicholas B"unger, Benedikt Herzog, Cindy Guerrero Toro, Benjamin Vella, Matthias Biner, Rico Br"utsch, Carmen Castroviejo Fernandez, Felix "Ottl, Norman Juchler, Armando Hoch, Bettina Hochreiter, Sven Hirsch, Sebastiano Caprara
Published: 8/4/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2608.00195v1 Announce Type: cross Abstract: High-resolution 3D segmentation of hip and shoulder anatomy from CT and MRI is essential for surgical planning, yet frozen segmentation models often fail under domain shift. CNN-based expert models are fully automatic but lack adaptability, whereas p...
192. TRACE-TS: Attribution-Grounded and Traceable Sensor-Language Reasoning for Human Activity Understanding ​
Author: Sparsh Rastogi, Tanmay Kumar, Baiyu Chen, Jatin Bedi, Zechen Li, Flora D. Salim
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.ET, cs.LG
arXiv:2608.00200v1 Announce Type: cross Abstract: Wearable sensors capture fine-grained motion patterns that support rich behavioral understanding, yet most existing methods reduce these signals to activity labels. Recent LM-based approaches generate natural-language explanations for sensor data, bu...
193. A reproducible and extensible framework for benchmarking competing risks survival models ​
Author: Bego~na B. Sierra, Colin McLean, Peter S. Hall, Sarah Friedrich-Welz, Catalina A. Vallejos
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.00271v1 Announce Type: cross Abstract: A wide range of statistical and machine learning methods have been proposed for survival analysis with competing risks, where the occurrence of one event (i.e., cancer death) precludes the occurrence of other events (i.e., cardiovascular disease deat...
194. Towards General Language-Conditioned Latent Safety Filters ​
Author: Ihab Tabbara, Yuxuan Yang, Hussein Sibai
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SY, eess.SY
arXiv:2608.00315v1 Announce Type: cross Abstract: Robot policies are becoming increasingly general, with vision-language-action (VLA) models enabling a single policy to execute diverse tasks specified in natural language. Safe deployment, however, requires adapting not only to new tasks but also to ...
195. RMSWeb: Reflection, Failure-Mode Mining, and Salvage-DS for Web Agent Reinforcement Learning ​
Author: Chengbo Liu, Lifang Zhou, Ruijie Yan, Pei Tan, Ao Sun, Haojun Huang, Guichun Hua, Sining Wei, Yining Chen, Yingying He, Yutao Xie
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.00335v1 Announce Type: cross Abstract: Compact web agents can reduce deployment cost, but training them poses challenges in both data collection and post-SFT reinforcement learning (RL). Successful trajectories are expensive to collect and often contain inefficient detours. After supervis...
196. CurveShift: Is Agent Progress Scalar? Separating Level from Shape ​
Author: Hanwen Xing, Pengyun Wang, BingXu Meng, Kumail Alhamoud, Xiang Li, Jicheng Wang, Xin Yu, Xinyang Han, Xiaomin Li, Philip Torr, Yuexing Hao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.00355v1 Announce Type: cross Abstract: Progress in large language models is often summarized using a single scalar measure, such as a time horizon, a latent ability estimate, or an aggregate benchmark score. These summaries capture the overall performance, but they do not test whether pro...
197. Pretrain on Small Synthetic Data, Scale Large for Free: Symmetry-Aware Foundation Model for Logic Rule Induction ​
Author: Yin Jun Phua
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LO, cs.AI, cs.LG
arXiv:2608.00383v1 Announce Type: cross Abstract: Logical rule induction seeks interpretable rules that transfer across propositional schemas. This requires respecting symmetries: atom naming, example order, polarity flips, and label swap. Enforcing exact symmetry by construction lets one trained in...
198. LOCUS-DT: Localization via Observation-Conditioned Uncertainty Scoring with Digital Twins ​
Author: Haozhe Lei, Roberto Bomfin, Marwa Chafii, Sundeep Rangan
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.LG, cs.RO
arXiv:2608.00406v1 Announce Type: cross Abstract: Accurate indoor localization is essential for emerging applications in robotic navigation and search and rescue. While classical methods typically focus on single-point estimates, complex indoor environments with heavy blockage and multipath propagat...
199. From Digital to Physical Reservoir Computing: Co-Optimizing Soft Robotic Reservoirs via Dynamics Matching ​
Author: Nicola Visentin, Maximilian St"olzle, Mariano Ram'irez Montero, Francesco Braghin, Daniela Rus, Cosimo Della Santina
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.00484v1 Announce Type: cross Abstract: Soft robotic substrates are promising for Physical Reservoir Computing (PRC) because their compliant nonlinear dynamics can provide temporal memory, high-dimensional state transformations, and efficient inference. However, physical reservoirs are oft...
200. The Bayesian Reflex: A Predictive Coding Engine for Artificial Intelligence ​
Author: Sourabh Bhattacharya
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.00492v1 Announce Type: cross Abstract: Predictive coding offers a powerful theory of cortical computation, but corresponding scalable algorithmic implementations for artificial intelligence have remained elusive. This paper introduces the Bayesian reflex, a computational framework that di...
201. Recursive Gaussian Processes and the Bayesian Brain ​
Author: Moumita Das, Dipanjan Ray, Sourabh Bhattacharya
Published: 8/4/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG, stat.ML
arXiv:2608.00503v1 Announce Type: cross Abstract: Predictive coding offers a powerful framework for cortical computation, yet scalable implementations that respect both Bayesian exactness and neurobiological constraints remain scarce. We bridge this gap by formally connecting predictive coding to Re...
202. RadYOLO: Computationally Efficient 3D Object Detection and Segmentation in CT and MRI ​
Author: Kai Geissler, Laurens M"uller-Groh, Hans Meine
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.00508v1 Announce Type: cross Abstract: Object detection and segmentation in three-dimensional medical images is a very active area of research. However, most proposed deep learning models carry a high computational cost, and only few aim to be broadly applicable, achieve high detection pe...
203. Native Multilingual Chain-of-Thought Reasoning in Low-Resource Southeast Asian Languages ​
Author: Sean Gip Lim, William Chandra Tjhi, Hai Leong Chieu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.00533v1 Announce Type: cross Abstract: Large Language Models have achieved substantial progress in reasoning capabilities. Yet in low-resource native settings, many suffer from cross-lingual collapse, reverting to English during intermediate steps that require complex logical reasoning. T...
204. UOT-IR: Structured Routing of High-Polyphony Symbolic Music into Fixed-Budget Representations ​
Author: Ziyue Kang, Nan Nan, Chenhao Lin, Xiaohong Guan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG
arXiv:2608.00576v1 Announce Type: cross Abstract: High-polyphony symbolic music is increasingly used in generation, analysis, and arrangement, yet many downstream tasks require bounded representations with fixed tracks or slots. Converting richly orchestrated scores into compact forms is therefore n...
205. A False Average: Chain-of-Thought Monitors Collapse Where They Are the Only Defense ​
Author: Shikhar Shiromani, Leo Richter
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG
arXiv:2608.00583v1 Announce Type: cross Abstract: Chain-of-thought (CoT) monitoring is meant to catch the reward hacks that look clean in the actions and betray themselves only in the reasoning. We show that this is exactly where an adversary who controls the reasoning can defeat it. Rewriting only ...
206. Element-Aware Group Learning for E-Commerce Image Generation ​
Author: Jingtong Chen, Jiahui Wang, Xue Zhao, ShaoGuo Liu, Minghao Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.00584v1 Announce Type: cross Abstract: Recent advances in image generation and editing have made prompt quality a key bottleneck for e-commerce creatives. Vision-language models (VLMs) can generate image-editing prompts from product images and metadata, but further improving their prompt-...
207. Verification Without Sufficiency: Per-Chunk Filtering Fails on Multi-Hop RAG, and Decomposition Repairs It ​
Author: Randhir Kumar
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.IR, cs.LG
arXiv:2608.00585v1 Announce Type: cross Abstract: Verification for retrieval-augmented generation usually scores each retrieved chunk and drops the ones that fail. We show this cannot work for multi-hop questions, and show what does. Per-chunk scoring assumes one chunk is a sufficient premise for th...
208. Uncertainty-guided active learning for surrogate prediction of stream-finishing wear fields ​
Author: Anand Kumar, Puli Saikiran, Vineet Dawara, Koushik Viswanathan
Published: 8/4/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG, physics.data-an
arXiv:2608.00593v1 Announce Type: cross Abstract: In stream finishing, the wear experienced by a workpiece depends strongly on its orientation within the rotating abrasive media. Determining suitable orientations to achieve uniform wear requires evaluating the wear-rate field over all feasible orien...
209. Beyond Lanes: Traffic Flow Dynamics in Disordered Conditions Based on High-Resolution Trajectory Data ​
Author: Shrey Agrawal, Gowri Asaithambi, Venkatesan Kanagaraj, Martin Treiber, Ostap Okhrin, Harish Babu Kumara
Published: 8/4/2026, 4:00:00 AM
Categories: physics.soc-ph, cs.LG
arXiv:2608.00602v1 Announce Type: cross Abstract: Disordered traffic flow is characterized by weak or non-existent lane discipline in the presence of strong vehicle heterogeneity and continuous lateral interactions, challenging traditional lane-based modeling assumptions. This study presents an empi...
210. Simulation-Based Plate-Reverb Parameter Estimation from a Single Impulse Response ​
Author: Minhui Lu, Joshua D. Reiss
Published: 8/4/2026, 4:00:00 AM
Categories: eess.AS, cs.LG
arXiv:2608.00656v1 Announce Type: cross Abstract: We present a simulation-trained, non-iterative estimator for Task A of the 1st DAFx Parameter Estimation Challenge. Each unnormalized plate-reverb impulse response is summarized by amplitude, spectral, and decay descriptors, and an ensemble of tree r...
211. Causal Inference with Unstructured Treatments ​
Author: Kevin Christian Wibisono, Yixin Wang
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.00657v1 Announce Type: cross Abstract: Causal inference usually concerns a scalar treatment, yet in many problems the treatment is unstructured: a text, an image, or a sequence of clinical decisions. Consider an instructor writing a course description to attract more students: the treatme...
212. Band-Count Dense Modal Estimation with Fixed-Frequency Differentiable Resonator Refinement ​
Author: Minhui Lu, Joshua D. Reiss
Published: 8/4/2026, 4:00:00 AM
Categories: eess.AS, cs.LG
arXiv:2608.00667v1 Announce Type: cross Abstract: Task B of the 1st DAFx Parameter Estimation Challenge requires estimating the frequencies, decay rates, gains, and number of modes in a dense plate-reverb impulse response. Weak and overlapping modes make sparse peak detection prone to severe underco...
213. Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors ​
Author: Alexander Scheinker
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, physics.comp-ph, physics.plasm-ph
arXiv:2608.00675v1 Announce Type: cross Abstract: Autoregressive models accumulate error over long rollouts, yet at deployment there is no ground truth to measure it against. We train a single conditional latent diffusion model that steps a dynamical system forward or backward in time via a directio...
214. Evolutionary Curriculum Learning Improves Biological Sequence Modeling ​
Author: Richard Zhu, Kento Nishi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-bio.BM, stat.ML
arXiv:2608.00697v1 Announce Type: cross Abstract: Variational autoencoders (VAEs) trained on multiple sequence alignments (MSAs) have emerged as powerful generative models for biological sequences, with applications ranging from disease variant prediction to functional RNA design. However, standard ...
215. Augmented Inverse Hybrid Weighting: Robust Inference under Deterministic and Random Distribution Shifts ​
Author: Ying Jin, Ying Jin, Dominik Rothenh"ausler
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, math.ST, stat.TH
arXiv:2608.00701v1 Announce Type: cross Abstract: Reweighting source samples to match a target covariate distribution is a standard response to distribution shift when generalizing evidence from one population to another. This strategy is well suited to deterministic, learnable covariate discrepanci...
216. Staged Multi-Agent Training (SMAT) for Hip Exoskeletons: Metabolic and Biomechanical Validation of a Simulation-Trained Co-Adaptive Controller ​
Author: Yifei Yuan, Jakob Wolf, Ghaith Androwis, Xianlian Zhou
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.00715v1 Announce Type: cross Abstract: Learning-based controllers can deliver exoskeleton assistance after training entirely in physics-based simulation, yet few controllers that address human-device co-adaptation have been validated on real users by whole-body metabolic measurement, the ...
217. Generated Images Are Easier to Forget: A Machine Unlearning Perspective for Synthetic Image Detection ​
Author: Jun Nie, Yonggang Zhang, Tongliang Liu, Yiu-ming Cheung, Bo Han, Xinmei Tian
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.00716v1 Announce Type: cross Abstract: Robust detection of generated images is critical to counter the misuse of generative models. Existing methods primarily depend on learning from human-annotated training datasets, limiting their generalization to unseen distributions. In contrast, lar...
218. CascadeLUT: Information-Ordered Streaming Inference for Bandwidth-Constrained FPGAs ​
Author: Oliver Cassidy, Marta Andronic, George A. Constantinides
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2608.00720v1 Announce Type: cross Abstract: Mapping neural networks to FPGAs enables low-latency, energy-efficient inference, particularly for lookup table (LUT)-based models that eliminate multipliers and map directly to reconfigurable fabric. While prior work achieves high compute efficiency...
219. Experience-Calibrated Contrastive Decoding for Mitigating Hallucinations in LM-Based Text-to-Speech ​
Author: Chenlin Liu, Minghui Fang, Zhonghao Bi, Zekai Su, Rong Wang, Jiqing Han
Published: 8/4/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, eess.SP
arXiv:2608.00722v1 Announce Type: cross Abstract: Language model-based text-to-speech (LM-based TTS) remains vulnerable to speech hallucinations that deviate from the target text. Existing mitigation mainly relies on architectural changes or additional training, while decoding-time control remains u...
220. An Uncertainty-Driven Hybrid Deep Learning Approach for Broad-Coverage RF Modulation Recognition ​
Author: Nurettin Safak, Durdu Can Yerdeyatar, Muhammet Sefa Demirel, Alperen Marasli, Taha Eren Atmaca, Ozgun Ersoy
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.00796v1 Announce Type: cross Abstract: Automatic RF modulation recognition is of critical importance in spectrum monitoring, electronic warfare, and cognitive radio applications, where low signal-to-noise ratio (SNR) conditions and the growing diversity of modulation schemes limit the per...
221. SoniSpeech: A Large-Scale Open-Vocabulary Tri-Modal Dataset for Wearable Silent Speech Interfaces ​
Author: Ruidong Zhang, Jiacheng Liu, Fran\c{c}ois Guimbreti`ere, Cheng Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.SD, cs.HC, cs.LG
arXiv:2608.00803v1 Announce Type: cross Abstract: Wearable silent speech interfaces (SSIs) are limited to small, closed vocabularies. Approaches achieving larger vocabularies require obtrusive hardware such as facial electrodes. We present SoniSpeech, the first large-scale, open-vocabulary, trimodal...
222. Pruned BPE: Post-training Visibility Pruning and Token Reallocation for Byte Pair Encoding ​
Author: Kenny Shao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.00837v2 Announce Type: cross Abstract: Byte Pair Encoding (BPE) is widely used for subword tokenization, but standard BPE exposes every learned merge token to the downstream model, including tokens that mainly serve as intermediate construction units and rarely appear in the final encoded...
223. Partially-Observable Transmission Control for UAV-Enabled Federated Learning in IoT Networks ​
Author: Masoud Ghazikor, Zhou Ni, Morteza Hashemi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT
arXiv:2608.00855v1 Announce Type: cross Abstract: Uncrewed aerial vehicle (UAV)-enabled federated learning (FL) can provide flexible, on-demand edge intelligence for large-scale IoT deployments, but operating in shared unlicensed bands makes uplink update delivery interference-coupled and unreliable...
224. Explainable Hybrid Feature Selection for Intrusion Detection in Internet of Medical Things Environments ​
Author: Amira Berrezzek, Hayet Djellali, Giulio Mallardi, Lamia Mahnane
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.00869v1 Announce Type: cross Abstract: Internet of Medical Things (IoMT) networks are hard to protect: devices are heterogeneous, computing resources are scarce, and traffic must be analyzed in real time. We present an intrusion detection system that addresses these constraints through fe...
225. PhenoStitch: Training-Free Panoptic Crop Mapping from Satellite Image Time Series ​
Author: Xuechen Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.00870v1 Announce Type: cross Abstract: Panoptic crop mapping requires both delineating individual agricultural parcels and assigning a crop type to each parcel from satellite image time series. Existing approaches typically rely on dense parcel-level annotations and task-specific model tr...
226. A Sequence-to-Sequence ConvLSTM Approach for Leaf Area Index Forecasting over the South-Central United States ​
Author: Zhixing Ruan, Lixin Lu
Published: 8/4/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG
arXiv:2608.00879v1 Announce Type: cross Abstract: Leaf Area Index (LAI) is a fundamental biophysical variable governing land-atmosphere interactions; however, LAI forecasting at high spatial resolution remains an unsolved challenge. While recent machine learning approaches have demonstrated LAI esti...
227. Learning Not to Optimize: Physics-Informed Action-Space Reshaping for Intent-Based Network Control ​
Author: Zuyuan Zhang, Vaneet Aggarwal, Tian Lan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.NI, cs.LG
arXiv:2608.00908v1 Announce Type: cross Abstract: Modern network policy control maps intent to sequential placement-control decisions. Bellman-style policy optimization primarily asks which action to optimize, while constraints are commonly handled through penalty, barrier, or Lagrangian mechanisms....
228. Tevatron Meets Megatron: Expert-Parallel LLM Reranker Training on an Academic Budget ​
Author: Zhichao Xu, Xueguang Ma, Shengyao Zhuang, Luyu Gao, Wenqian Ye, Yu Wang, Jamie Callan, Jimmy Lin
Published: 8/4/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.00916v1 Announce Type: cross Abstract: Modern reranking recipes---billion-scale cross-encoders, mixture-of-experts (MoE) backbones, and distillation against strong teachers---have outpaced the training infrastructure available to most academic groups. Existing Tevatron reranker training r...
229. Rethinking PPG-based Sleep Staging: Datasets, Metrics, and Benchmarks ​
Author: Shuntian Zheng, Jiawei Wang, Cong Fu, Huan Yu, Chen Chen, Yu Guan, Sai Gu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.HC, cs.LG
arXiv:2608.00943v1 Announce Type: cross Abstract: Automated sleep staging assigns discrete stage labels to successive time epochs throughout an overnight recording; conventionally each window spans at least 30 seconds, reflecting the minimum temporal resolution of the clinical scoring standard. Wear...
230. GraRe: Grasp Candidate Re-Ranking for Frozen 6-DoF Grasp Detectors ​
Author: Jibao Yuan, Yuhui Zhao, Yinzhen Lv, Chao Xu, Shun Li, Chenxi Deng, Shaofei Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG
arXiv:2608.00946v1 Announce Type: cross Abstract: Existing 6-DoF grasp detectors typically rank grasp candidates by detector confidence. However, our analysis on GraspNet-1Billion shows that detector confidence is often poorly aligned with grasp quality, causing successful grasp candidates to be ran...
231. Model-Agnostic FDR Control via Group Gaussian Mirror and Permutation SHAP ​
Author: Jiaan Han, Junxiao Chen, Yanzhe Fu
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.00989v1 Announce Type: cross Abstract: Most FDR-controlled feature selection methods are designed for coordinate-wise hypotheses, where each feature has a single weight or importance score. This abstraction fails in sequential and grouped models, where one original feature is represented ...
232. Using Lower-Bound Representations for Trajectory Similarity Learning ​
Author: Liwei Deng, Haotian Meng, Yupu Zhang, Yan Zhao, Torben Bach Pedersen, Kai Zheng, Christian S. Jensen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DB, cs.LG
arXiv:2608.01039v1 Announce Type: cross Abstract: Trajectory similarity learning is fundamental to efficient trajectory retrieval under complex distance measures. Existing learning-based methods typically rely on embeddings trained to approximate trajectory distances or rankings, but they often lack...
233. On the Limits of Machine-Learned Ranking for Modern Microarchitectural Policies ​
Author: Yanxin Zhang, Shayne Wadle, Yuxuan Xiong, Zheyu Fu, Trivikram Krishnamurthy, Karu Sankaralingam
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2608.01041v1 Announce Type: cross Abstract: Machine-learning predictors estimate processor performance far faster than cycle-level simulation. For design-space exploration, however, the valuable test is not merely reproducing the usual hardware ordering, but identifying how different hardware ...
234. What Could the Agent See at 19:05? Generating Temporal Enterprise Scenarios from Real Research and Replaying Them to Evaluate Agents ​
Author: Tezan Sahu, Himani Arora
Published: 8/4/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.HC, cs.LG
arXiv:2608.01042v1 Announce Type: cross Abstract: Enterprise AI agents act across many apps whose data changes continuously, so an answer is correct only relative to what data existed and who could see it at the moment it was asked. Offline evaluation today grades against a single static snapshot, e...
235. FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds ​
Author: Kapil Wanaskar, Gaytri Jena, Aman Chadha, Vinija Jain, Vasu Sharma, Amitava Das
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG
arXiv:2608.01049v1 Announce Type: cross Abstract: World models have attracted significant attention for their ability to capture and predict the structure and dynamics of the physical world. In this emerging landscape, Joint Embedding Predictive Architectures (JEPA) offer a particularly compelling d...
236. When Collaboration Becomes a Trigger: Collective Evidence-Threshold Backdoors in Multi-Agent Systems ​
Author: Jia-Hao Xiao, Lei Feng, Min-Ling Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2608.01085v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) extend LLM capabilities through iterative communication and shared contexts. However, this collaboration introduces a vulnerability: backdoor behavior can be activated when peer evidence reaches a hidden threshold,...
237. MA-HEAD-Net: Adaptive Rule-Guided Multi-Agent DRL for AoI Minimization in UAV-Assisted Emergency Networks ​
Author: Yixin Zhang, Zhuohui Yao, Wenchi Cheng, Walid Saad
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.01128v1 Announce Type: cross Abstract: In post-disaster scenarios, unmanned aerial vehicles (UAVs) are critical for establishing emergency communication networks. For time-critical rescue missions, information freshness is crucial because decisions based on outdated data may lead to ineff...
238. Learning-Based Stochastic Optimal Control with Infinite-Horizon Probabilistic Constraints ​
Author: Francesco Cordiano, Kanghui He, Bart De Schutter
Published: 8/4/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.SY, eess.SY
arXiv:2608.01151v1 Announce Type: cross Abstract: In this paper, we consider stochastic optimal control problems with infinite-horizon joint chance constraints. By means of an appropriate state augmentation, we reformulate the original problem as a constrained Markov decision process, in which both ...
239. 3DZip: Spatial-Aware Feature Diversity-Guided Token Compression for 3D Question Answering ​
Author: Changwoo Baek, Kyeongbo Kong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.01185v1 Announce Type: cross Abstract: Recent 3D vision-language models (3D VLMs) construct geometry aware tokens by projecting 2D visual features into world coordinates, enabling spatial reasoning for tasks such as 3D question answering. However, this design generates thousands of tokens...
240. Humans Are More Diverse: Frontier LLMs Show Extreme Policies in Idealised AI Development Races ​
Author: Phu Hoa Pham, Duy Minh Dao Sy, Trung Kiet Huynh, Phu Quy Nguyen Lam, Chi Nguyen Tran, Minh Trung Le, Phong Hao Le, Dinh Nam Nguyen, Thien Ky Nguyen Dong, Elias Fernandez Domingos, Le Hong Trang, The Anh Han
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.GT, cs.LG, cs.MA
arXiv:2608.01193v1 Announce Type: cross Abstract: An AI development race creates a multi-agent safety dilemma. Each company can develop slowly and safely, or move faster while taking a risk that may remove its final reward. We use this repeated game to study strategic safety behaviour among large la...
241. Hybrid Quantum Neural Networks: Theory, Implementations, and Applications ​
Author: L'eo Monbroussou, Maniraman Periyasamy, Viacheslav Kuzmin, Pavel Sekatski, Viktoria Patapovich, Asel Sagingalieva, Alexey Melnikov
Published: 8/4/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG
arXiv:2608.01194v1 Announce Type: cross Abstract: Artificial intelligence has been transformed by deep neural networks, yet the search for new learning architectures continues. Quantum machine learning offers one such direction, and hybrid quantum neural networks, which combine classical neural-netw...
242. Fruit-HSNet: A Machine Learning Approach for Hyperspectral Image-Based Fruit Ripeness Prediction ​
Author: Ahmed Baha Ben Jmaa, Faten Chaieb, Anna Fabija'nska
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.01202v1 Announce Type: cross Abstract: Fruit ripeness prediction (FRP) is a classification-based agricultural computer vision task that has attracted much attention, thanks to its wide-ranging advantages in agriculture field for both pre-harvest and post-harvest management. Accurate and t...
243. Climate-Dyna Deep Hedging for XVAs: Model-Based Reinforcement Learning, Residual Climate HVA, and Hedge-Instrument Discovery ​
Author: Xiaozhen Wang, Francois Buet-Golfouse
Published: 8/4/2026, 4:00:00 AM
Categories: q-fin.MF, cs.LG, q-fin.RM
arXiv:2608.01208v1 Announce Type: cross Abstract: For a trading desk, residual climate hedging valuation adjustment (HVA) is the climate cost left after its inherited hedge and any admissible overlay have been taken into account; it therefore cannot be inferred from a stand-alone stress loss. We obt...
244. Amortizing the Calibration Triple: A Projection-Consistent Neural Operator for Local-Stochastic Volatility ​
Author: Xiaozhen Wang, Ana"is Despr'es, Martin Dureau, Francois Buet-Golfouse
Published: 8/4/2026, 4:00:00 AM
Categories: q-fin.MF, cs.LG
arXiv:2608.01217v1 Announce Type: cross Abstract: Local-stochastic volatility (LSV) combines vanilla marginals with richer smile dynamics, but calibration requires a slow, noisy and sequential McKean--Vlasov fixed point. We learn a projection-consistent operator for the calibration triple. Given fin...
245. Using Non-Lipschitz Signum-based Functions for Distributed Optimization and Machine Learning: Trade-off Between Con-vergence Rate and Optimality Gap ​
Author: Mohammadreza Doostmohammadian, Amir Ahmad Ghods, Alireza Aghasi, Zulfiya R. Gabidullina, Hamid R. Rabiee
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, eess.SP, math.OC
arXiv:2608.01220v1 Announce Type: cross Abstract: In recent years, the prevalence of large-scale data-sets and the demand for sophisti-cated learning models have necessitated the development of efficient distributed ma-chine learning (ML) solutions. Convergence speed is a critical factor influencing...
246. RestoreKV: Recovering Full-Cache Behavior Under Aggressive Query-Agnostic KV Cache Eviction ​
Author: Changwoo Baek, Seungjun Shin, Kyeongbo Kong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.01247v1 Announce Type: cross Abstract: Query-agnostic KV cache eviction compresses a context once and reuses the resulting cache for arbitrary future queries, but performance can collapse under tight budgets. Existing methods primarily improve which original KV pairs are retained. We intr...
247. How fine a change can moments see? A scale law for detecting distribution shift, with a kernel calibration rule ​
Author: Adel Kaleche
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.01268v1 Announce Type: cross Abstract: Detecting that a stream of high-dimensional embeddings has changed is usually framed as a choice of statistic. We give a scale law that constrains any moment-based choice and test it against topological alternatives. The law: certifying a feature of ...
248. Latent Softmax for Data-Efficient Phoneme-Based Multilingual ASR Across Tonal and Non-Tonal Languages ​
Author: Saierdaer Yusuyin, Nanling Jiang, Hao Huang, Zhijian Ou
Published: 8/4/2026, 4:00:00 AM
Categories: eess.AS, cs.LG
arXiv:2608.01281v1 Announce Type: cross Abstract: Phoneme-based multilingual automatic speech recognition (ASR) can share acoustic evidence across languages more directly than language-specific subword modeling. When tonal and non-tonal languages are jointly trained, however, their supervision granu...
249. Active Regression for Single-Index Models with Unknown Link Functions ​
Author: Chansophea Wathanak In, Yi Li, Wai Ming Tai, Xuan Wu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2608.01287v1 Announce Type: cross Abstract: This paper studies active regression for single-index models under general $\ell_p$-loss with an unknown $1$-Lipschitz link function $f$, formulated as $\min_{f,x} |f(Ax)-b|_p^p$ with full access to $A$ but coordinate-query access to $b$. Prior wor...
250. UDT: Reconciling U-Nets and Diffusion Transformers with Data-Adaptive Token Reduction ​
Author: Junno Yun, Ya\c{s}ar Utku Al\c{c}alar, Mehmet Ak\c{c}akaya
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.01298v1 Announce Type: cross Abstract: Diffusion Transformers (DiTs) have emerged as a core architecture in generative modeling due to their scalability and adaptability to multimodal tasks. DiTs comprise isotropic transformer blocks, and learn representations progressively across depth, ...
251. Sheaf-theoretic Signal Processing on Graphs: Spectral Theory, Filtering, and Sampling ​
Author: Gabriele D'Acunto, Leonardo Di Nino, Paolo Di Lorenzo, Sergio Barbarossa
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG
arXiv:2608.01318v1 Announce Type: cross Abstract: Modern sensing, communication, and learning systems generate heterogeneous network signals, with local data differing in dimension, modality, and geometric structure. Processing such data requires a mathematical framework capable of simultaneously mo...
252. Dense Language Generation Made Simple: Deterministic, Randomized, and Multi-Order Algorithms ​
Author: Ziyi Cai, Shuangping Li, Yiheng Shen, Kangning Wang, Peng Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DS, cs.AI, cs.CL, cs.DM, cs.LG
arXiv:2608.01320v1 Announce Type: cross Abstract: Language generation in the limit is a theoretical framework for studying how a generator can learn to produce new valid strings from a stream of positive examples. In this model, an adversary chooses an unknown language from a countable family and en...
253. Asleep at the Wheel: JEPA's Limitations in Evaluating Novel Driving Data ​
Author: Advait Pavuluri, Shamik Karkhanis, Uzma Mushtaque
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.01336v1 Announce Type: cross Abstract: Modern autonomous-driving fleets record far more video than human reviewers can inspect. This motivates the need for an automatic clip triage mechanism, to surface rare and review-worthy clips, so that driving models can be fine-tuned to better handl...
254. Why Formal Monitors Fail: Attack Distribution Entropy as a Coverage Bound for LTL-Based LLM Agent Safety ​
Author: Ruiyang Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2608.01388v1 Announce Type: cross Abstract: Runtime safety monitors based on Linear Temporal Logic (LTL) and finite automata (FSA) are increasingly deployed to intercept unsafe tool-call sequences in LLM agents. Yet the same monitor achieves 68-75% attack coverage on some model architectures a...
255. Reusing Rollouts under Policy Lag: Prefix-Normalized Policy Optimization for LLM Reinforcement Learning ​
Author: Wenhao Zhang, Yibo Xie, Rui Wang, Jiahua Yang, Lei Jiang, Zibo Yang, Yawei Wang, Jiali Xu, jasperawang, Haoyang Long, Huan Xiong, alantzhao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.01418v1 Announce Type: cross Abstract: Autoregressive rollout generation is a major computational cost in reinforcement learning for large language models. Reusing each rollout batch for additional learner updates amortizes this cost, but later updates become increasingly off-policy as th...
256. QR-Erase: Efficient Subspace-Based Machine Unlearning with Layer Localization ​
Author: Tyler Lizzo, Larry Heck
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.01422v1 Announce Type: cross Abstract: Machine unlearning seeks to remove targeted information from trained models without requiring costly retraining. Existing optimization-based methods often degrade unrelated capabilities, while subspace-based approaches rely on computationally expensi...
257. Scoring Rules! Statistical and Strategic Alignment for Text Evaluation Metrics ​
Author: Shengwei Xu, Yuxuan Lu, Yifan Wu, Jason Hartline, Grant Schoenebeck
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.GT, cs.LG
arXiv:2608.01423v1 Announce Type: cross Abstract: Reference-based text evaluation metrics, which are widely used to assess natural language generation systems, score a candidate response by comparing it with a reference response. The reliability of an evaluation metric is usually judged by its stati...
258. Training Small LLMs as Spatial Multi-Agent Policies ​
Author: Yi Mao, Andrew Perrault
Published: 8/4/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2608.01425v1 Announce Type: cross Abstract: Training LLM-based multi-agent systems with multi-agent reinforcement learning is rapidly gaining traction, and a parallel line of work argues that such systems should be judged by their behavior, not only their reward. We take up both threads in spa...
259. When Replanning Becomes the Bottleneck: Budgeted Replanning for Embodied Agents ​
Author: Shuaijun Liu, Feiyang You, Xingwei Chen, Ningxin Su
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2608.01428v1 Announce Type: cross Abstract: Embodied agents replan frequently to recover from execution drift, partial observability, and coordination hazards, but each LLM-based replanning call can consume an accumulated textual context that grows over time and across agents. Once this contex...
260. PolymerGPT: Multi-property Optimization with a Decoder-Based GPT Model for Generative Polymer Design ​
Author: Charlie Pyle, Adarsh Gadari, C. Adrian Figg, Zhenquan Jia, Yaohang Li, Chunjiang Zhu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cond-mat.mtrl-sci, cs.CE, cs.LG
arXiv:2608.01431v1 Announce Type: cross Abstract: Polymer property prediction and inverse generative design targeting desired properties are two crucial tasks in machine learning-assisted polymer design. While the former has received considerable attention, there have been limited methods developed ...
261. DynamicManip: Enabling Dynamic Manipulation from a Single Static Demonstration ​
Author: Haoran Liao, Pengyue Wang, Shuoyu Chen, Kehan Cheng, Xuhang Chen, Yuhao Lin, Mu Lin, Zhizhao Liang, Xiaoyi Fan, Chengyi Xing, Dan Niu, Yi-Lin Wei, Wei-Shi Zheng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG
arXiv:2608.01452v1 Announce Type: cross Abstract: Dynamic manipulation is a critical capability for robots operating in complex and dynamic environments, where robots must interact with objects that are moving or require rapid adjustments. However, learning models for dynamic manipulation tasks face...
262. How Benchmarks and Evaluation Protocols Shape Conclusions in Provenance-Based Intrusion Detection ​
Author: Lorenzo Guerra, Thomas Chapuis, Guillaume Duc, Pavlo Mozharovskyi, Van-Tam Nguyen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.01454v1 Announce Type: cross Abstract: Provenance-based intrusion detection systems (PIDS) frequently report strong performance, but the conclusions drawn from these results can be highly sensitive to benchmarking choices and evaluation protocols. We investigate this dependency by re-eval...
263. Slot2Text: Object-Centric Visual Tokenization for Efficient and Spatially Traceable Surgical MLLMs ​
Author: Guiqiu Liao, Matjaz Jogan, Daniel A. Hashimoto
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG
arXiv:2608.01473v1 Announce Type: cross Abstract: Multimodal large language models (MLLM) for surgical scene understanding typically inject hundreds of dense visual tokens into a language model, leading to costly inference and limited spatial traceability for generated answers. We present Slot2Text,...
264. Rapid Embodiment Adaptation for Quadrupedal Locomotion ​
Author: Dichen Li, Bo Ai, Nico Bohlinger, Jan Peters, Hao Su, Henrik I. Christensen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2608.01506v1 Announce Type: cross Abstract: Humans readily adapt their movements as their bodies change through aging, injury, or load carrying, but learning-based robot policies often break when hardware properties shift. We introduce an online embodiment adaptation framework for quadrupedal ...
265. Celty: SpMspV GPU Kernel and SIMT Co-Design for Efficient Dual-Sparse LLM Inference ​
Author: Ruokai Yin, Priyadarshini Panda
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2608.01536v1 Announce Type: cross Abstract: Large Language Models (LLMs) increasingly rely on sparsity to reduce inference cost, but most prior work targets a single sparsity source-either weight or activation-and optimizes for batched multi-user inference. Dual-sparsity, which combines unstru...
266. Dominant Arm Identification with Mixing and Recycling Observed Samples ​
Author: Jonghyun Sim, Wonyoung Kim
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.01545v1 Announce Type: cross Abstract: We study the problem of identifying the dominant arm in multi-armed bandits, where the objective is to find the action with the highest probability of exceeding the realized rewards of all other actions. Conventional mean-based and pairwise compariso...
267. Finite-Probe Total-Variation Certificates for Finite-Basis Drifting Models ​
Author: Sam Andersson, Ricky Mol'en
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT
arXiv:2608.01547v1 Announce Type: cross Abstract: Drifting objectives compare a target and model distribution through a vector field observed noisily at finitely many locations. We ask what distributional conclusion such a frozen measurement system warrants. For integrable antisymmetric interactions...
268. Emergence Invariance: From Symbolized Thought to Interface Refinement ​
Author: Yi Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.01548v1 Announce Type: cross Abstract: Language can be viewed as a formalized subset of thought: a consequence-governed symbolic structure projected from wider situated cognition. Large language models trained at scale exhibit compensatory emergence: sparse architectural primitives suppor...
269. Generalized Quadratic Gradient: A New Direction in Optimization via the Fusion of Positive-Definite Curvature Matrices and Gradients into A Unified Framework ​
Author: John Chiang
Published: 8/4/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2608.01552v1 Announce Type: cross Abstract: Quadratic Gradient (QG) is a Newton-type optimization framework that bridges first-order gradient descent and second-order optimization by incorporating curvature information into gradient updates. Simplified Quadratic Gradient (SQG) reduces the comp...
270. Does the Competitive Component of Adversarial Self-Play Improve Legal Reasoning? A Controlled Negative Result ​
Author: Miseog Shawn Kim
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.01559v1 Announce Type: cross Abstract: Adversarial self-play is an appealing recipe for legal reasoning: have a student model draft an argument, have an adversary attack it, and reward the student when its argument survives the attack. We designed exactly such a training signal -- a verif...
271. LieStoNet: Learning Lie Symmetries from Spatiotemporal Data for Stochastic Dynamical Systems ​
Author: Shida Liu, Abhishek Gupta, Sumit Sinha, L. Mahadevan
Published: 8/4/2026, 4:00:00 AM
Categories: cond-mat.stat-mech, cond-mat.dis-nn, cs.LG, math-ph, math.MP
arXiv:2608.01582v1 Announce Type: cross Abstract: Symmetry is central to modern machine learning and physics: invariances and equivariances improve sample efficiency, robustness, and out-of-distribution generalization, while symmetry principles guide scientific modeling. Yet for stochastic dynamical...
272. Semantic Alignment of AI Models: Concept Collapse, Checkpoint Dynamics, and Cross-Lingual Transfer ​
Author: Tyler Ashoff, Jordan Rodu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.01585v1 Announce Type: cross Abstract: Language model benchmarking is a difficult task. Outcome reasoning alone does not test the model's conceptualization of language and popular open-source benchmarks are quickly saturated or ingested as training data. It is important to test the model'...
273. Statistical comparisons of time-series feature sets on classification tasks ​
Author: Trent Henderson, Ben D. Fulcher
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML
arXiv:2608.01586v1 Announce Type: cross Abstract: In recent years, numerous open-source software libraries have been developed for computing sets of features from univariate time series. The type and number of features vary across these feature sets, which have been constructed with varying discipli...
274. The Label Defines the Timescale: Trait-State Limits of Temporal-Aggregate Learning ​
Author: Xizhe Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2608.01587v1 Announce Type: cross Abstract: Machine-learning benchmarks often pair a label that aggregates a long temporal horizon with input observed through one or a few short windows. Their apparent performance ceiling may therefore be an acquisition-protocol ceiling rather than a model-cap...
275. Thermalizing Stochastic Programs ​
Author: Mirko Amico, Andra\v{z} Jelin\v{c}i\v{c}, Colin Oscar Nancarrow, Leo Tyrpak, David Roberts, Seth Morton, Dalton Sakthivadivel, Ashwin Gopal, Guillaume Verdon
Published: 8/4/2026, 4:00:00 AM
Categories: cs.ET, cs.LG
arXiv:2608.01615v1 Announce Type: cross Abstract: We present a set of tools for mapping general stochastic programs to thermodynamic hardware designed for energy-efficient stochastic sampling. Given a target stochastic program expressed as a Directed Factor Graph (DFG) of stochastic channels, or equ...
276. Not the Dimension, the Norm: What Matters in Gradient-Free Weight Perturbation of Language Models ​
Author: Taeyeong Kim, Ahhyun Kim, TaeHyeon Kim, Unggi Lee
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.01624v1 Announce Type: cross Abstract: Adapting a language model to a task no longer requires training all of its weights, and a line of parameter-efficient methods has driven the trainable count from billions down to a handful of scalars. Gradient-free adaptation, which samples random we...
277. Bole: Efficient Tree Speculation for Hybrid-Attention Language Models ​
Author: Li Wang, Yi Su, Xiabao Wu, Chiran You, Yongchao Liu, Zhan Qiu, Juelu Zhang, Jiajun Zheng, Fangxin Liu, Jie Zhang, Chen Tian, Chengying Huan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DC, cs.CL, cs.LG
arXiv:2608.01651v1 Announce Type: cross Abstract: Hybrid-attention large language models combine full attention with recurrent linear attention to reduce long-context inference costs, yet their autoregressive decoding remains memory-bound. Tree speculative decoding offers an attractive acceleration ...
278. Non-KKT Accumulation in Entropic Mirror Descent ​
Author: Kuangyu Ding, Kim-Chuan Toh
Published: 8/4/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DS
arXiv:2608.01658v1 Announce Type: cross Abstract: For mirror descent generated by a Legendre kernel, perhaps one of the most basic question in optimization is this: must every accumulation point of a bounded mirror descent sequence be Karush--Kuhn--Tucker (KKT) stationary under proper stepsizes? We ...
279. LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing ​
Author: Wen Zan, Jiaqi Zhang, Jianchao Tan, Hong Liu, Cunguang Wang, Xiang Li, Duyue Ma, Guanyu Wu, Yifan Lu, Fengcun Li, Yerui Sun, Peng Pei, Yuchen Xie, Xunliang Cai
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.DC, cs.LG
arXiv:2608.01662v2 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) enables efficient long-context modeling through its Lightning Indexer. However, practical deployment remains constrained by the indexer's expensive $O(L^2)$ scoring overhead and the hardware-inefficient, discontinuous ...
280. FAU at ImageCLEF 2026 Task on Multimodal Reasoning Robust Candidate Scoring and Concise Multilingual Visual Answering ​
Author: Mohamed Basem, Vincent Christlein
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.01664v1 Announce Type: cross Abstract: We present our ImageCLEF 2026 Multimodal Reasoning system for the Visual Multiple Choice Question Answering (Visual MCQ) and Visual Open Question Answering (Visual OpenQA) subtasks. The challenge requires reliable reasoning over multilingual educatio...
281. Learning What to Remember: Test-Time Training via Context Distillation ​
Author: Zixuan Wang, Xingyu Dang, Rui-Jie Zhu, Zixin Wen, Hengyu Fu, Wenhao Chai, Jason D. Lee
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.01672v1 Announce Type: cross Abstract: Effective long-context modeling is not merely about retaining more of the past, but about preserving the information that may prove relevant later. Test-time training (TTT) is an appealing approach that performs online parameter updates for long-cont...
282. Understanding Sparse Attention Selectivity in Long-Context Foundation Models via Counterfactual Evaluation ​
Author: Xingyu Ren, Youran Sun, Chugang Yi, Haizhao Yang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.01676v1 Announce Type: cross Abstract: Sparse attention is widely deployed in long-context serving stacks, yet no framework audits how discarding blocks changes the influence of specific content on model output. We first establish that the phenomenon is real and causal: Block Sparse Flash...
283. Generative Brownian Bridge Diffusion In Motion Space For Enhanced Myocardial Strain Analysis ​
Author: Rishov Paul, Frederick H. Epstein, Miaomiao Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.01677v1 Announce Type: cross Abstract: Myocardial strain analysis of cardiac magnetic resonance (CMR) images provides an important tool for evaluating cardiac function. However, current techniques require either human-adjusted post-processing with suboptimal regional accuracy, or speciali...
284. CENTILE: A Telemetry Foundation Model Evaluated by the Decisions It Drives ​
Author: Zifan Zhang, Zhichao Hou, Tingxiang Ji, Yuchen Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.NI, cs.LG
arXiv:2608.01725v1 Announce Type: cross Abstract: Modern computing and networking infrastructure emits telemetry continuously, yet operators convert it into decisions with a separate predictor per task, entity, and horizon. One generative model, pretrained once over an operator's own event streams, ...
285. Reassessing the Feasibility of PPG-Based Non-Invasive Blood Glucose Level Estimation ​
Author: Supraja Ramesh, Markus Neufeld, Michael K"uttner, Tobias R"oddiger, Michael Beigl
Published: 8/4/2026, 4:00:00 AM
Categories: cs.HC, cs.LG
arXiv:2608.01820v1 Announce Type: cross Abstract: Non-invasive blood glucose level (BGL) estimation from photoplethysmography (PPG) holds great promise for wearable health monitoring, but results across studies are hard to compare due to inconsistent datasets, data leakage, and non-standardized eval...
286. DAVET: Denoising-Aware Visual Evidence Trajectory Allocation for Diffusion Vision-Language Models ​
Author: Yongkang Zhou, Xiang Xia, Cheng Yan, Fan Xu, Wuyang Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.01821v1 Announce Type: cross Abstract: Diffusion vision-language models (dVLMs) iteratively denoise masked responses while conditioning each denoising step on visual evidence, making visual conditioning a substantial recurring inference cost. Unlike autoregressive decoding, diffusion gene...
287. Tunneling the Loss Landscape: Bypassing Memorization with Monte Carlo Parameter Swapping ​
Author: Lai Shun Chan, Xiaotian Zhang, Yue Shang, Ge Zhang, Entao Yang
Published: 8/4/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cs.LG
arXiv:2608.01833v1 Announce Type: cross Abstract: Grokking is a striking phenomenon in neural network training, where a model can undergo a prolonged period of pure memorization before abrupt generalization. While previous works have attempted to interpret it through classical machine learning mecha...
288. Probabilistic Deep Learning for Drought Forecasting: Role of Internal Climate Variability ​
Author: Henri Funk, Cornelia Gruber, G"oran Kauermann, Helmut K"uchenhoff, Magdalena Mittermeier
Published: 8/4/2026, 4:00:00 AM
Categories: stat.AP, cs.LG, physics.ao-ph
arXiv:2608.01864v1 Announce Type: cross Abstract: Predicting drought risk is essential for anticipating impacts on water resources, agriculture, ecosystems, and climate adaptation planning. Yet drought forecasts remain uncertain because variability can substantially alter regional precipitation and ...
289. ReasonCast: Towards Explainable Time Series Forecasting with Reasoning ​
Author: Seunghan Lee, Jun Seo, Jaehoon Lee, Junhyeok Kang, Sangjun Han, Sungdong Yoo, Minjae Kim, Tae Yoon Lim, Dongwan Kang, Hwanil Choi, Soonyoung Lee, Wonbin Ahn
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.01875v1 Announce Type: cross Abstract: Most time series (TS) models are specialized for a single task, either understanding (i.e., returning text answers about a TS) or generation (i.e., returning a numeric forecast). Only recently have unified models begun to handle the two within a sing...
290. SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models ​
Author: Jing Wu, Jianhua Wu, Jiayi Guan, Jiahong Chen, Jinghui Lu, Hangjun Ye, Bingzhao Gao, Long Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG
arXiv:2608.01899v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) perform well on commonsense reasoning tasks but struggle with visual spatial reasoning. Most existing solutions introduce extra 3D prior inputs or external spatial encoders, which increase complexity and degrade the unde...
291. Automatic Annotation of Ancient Greek Vowel Length ​
Author: Albin Th"orn Cleland, Eric Cullhed
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.01935v1 Announce Type: cross Abstract: Prior work in Ancient Greek NLP relies on corpora that do not disambiguate the phonemic vowel length of alpha, iota, and ypsilon, together known as the dichrona. Depending on lexeme, morphology, sandhi, syntax, and conventions of period, genre, and v...
292. Look Ahead Before You Distill: Future Trajectory Validation of Teacher Guidance for Agentic On-Policy Distillation ​
Author: Chishui Chen, Yaoyou Fan, Te Sun, Yi Yang, Chenghao Sun, Delin Mao, Hongbo Qiao, Zuowei Zhang, Junxi Wang, Chenxing Sun, Yangen Hu, Lu Pan, Xuyang Liu, Linfeng Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.01953v1 Announce Type: cross Abstract: On-policy distillation (OPD) provides teacher supervision on states visited by the student, reducing the distribution gap between training and inference. However, in multi-turn agentic tasks, student deviations may accumulate over time, gradually mov...
293. TELLER: Non-intrusive Cross-Layer Root-Cause Analysis for LLM Inference ​
Author: Ruilin Xu, Junyi Li, Pengfei Chen, Zongxuan Xie
Published: 8/4/2026, 4:00:00 AM
Categories: cs.SE, cs.CL, cs.LG, cs.PF
arXiv:2608.01975v1 Announce Type: cross Abstract: Large language model (LLM) inference has evolved from an offline workload into a continuously operated software service, yet root-cause analysis remains difficult because a single request spans the inference engine, Python/C++ backend, host CUDA APIs...
294. Detecting Nonproperness of Likelihood Equations ​
Author: Xiaoxian Tang, Bican Xia, Tianqi Zhao
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.SC
arXiv:2608.01976v1 Announce Type: cross Abstract: Given an algebraic statistical model, a challenging problem is classifying the data according to the number of positive critical points of the likelihood function. The positive critical points are the positive solutions to an algebraic system, say li...
295. Learning-Based Collaborative MEC for LLM Inference with Soft-Deadline Awareness via Transformer-Enhanced PPO ​
Author: Ngoc Hung Nguyen, Bjorn Landfeldt
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DC, cs.LG, cs.NI
arXiv:2608.02031v1 Announce Type: cross Abstract: This paper investigates collaborative mobile edge computing (MEC) servers for large language model (LLM) inference under soft deadline constraints. In this system, to improve the quality of service, computations are expected to be completed within th...
296. D\'ej\`a Cue: Localizing States in Object Histories via Vocabulary-Relative Coordinates ​
Author: Haofan Cao, Zhichao You, Yunkai Yang, Liang Guo, Jie Wang, Chongshou Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.MM
arXiv:2608.02044v1 Announce Type: cross Abstract: Tracking links observations of the same object through visual change, yet cannot by itself determine when the object is empty or filled, intact or cut. We formulate identity-conditioned state-moment retrieval: given a tracked-object history and alter...
297. Adaptive Reconstruction of Bosonic Quantum States ​
Author: Vasilisa Usova, Phila Rembold, Ian Yang, Marco Rossignolo, Simone Montangero, Samuele Tosatto, Gerhard Kirchmair
Published: 8/4/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.02049v1 Announce Type: cross Abstract: Bosonic quantum systems provide a hardware-efficient platform for quantum information processing but remain challenging to characterise due to their large Hilbert space and the high measurement cost of state tomography. Existing approaches estimate t...
298. TextNCA: Neural Cellular Automata for Language Modeling via Hierarchical Local Attention ​
Author: Avni Mittal, Avinash Anand, Ashutosh Kumar, Dikshant Kukreja, Kritarth Prasad, Sushane Dulloo, Erik Cambria, Timothy Liu, Zhengkui Wang, Rajiv Ratn Shah
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.02050v1 Announce Type: cross Abstract: Can a strictly local, iterated, weight-shared computation primitive support language modelling, and which of those three properties actually drives the model's behaviour? We define \textsc{TextNCA}, a 1D causal windowed-attention realisation of the N...
299. A Comparative Analysis of MLP and Kolmogorov-Arnold Networks (KAN) for Faster-than-Nyquist (FTN) Signaling Detection ​
Author: Sude Ertan, Osman Tokluoglu, Enver Cavus
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.02062v1 Announce Type: cross Abstract: Faster-than-Nyquist signaling improves spectral ef- ficiency by deliberately introducing inter-symbol interference. Classical sequence detectors such as BCJR can approach optimal performance, but their computational cost grows rapidly with channel me...
300. Open-DiffLoco: Open-Source Differentiable Learning for Deployable Blind Quadruped Locomotion ​
Author: Martin Opat
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SY, eess.SY
arXiv:2608.02069v1 Announce Type: cross Abstract: Developing deployable locomotion policies through conventional reinforcement learning often requires complex reward engineering and expensive training times. While differentiable simulation offers a highly efficient alternative, open-source tools cap...
301. STEAM:ASpatio-TEmporal Alignment Mixture-of-Experts Model with Hierarchical Pre-training for EEG Decoding ​
Author: Zhu Chen, Dingkun Liu, Yuheng Chen, Dongrui Wu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.02070v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) have been widely used in motor rehabilitation, disease diagnosis, and other neural engineering scenarios. However, conventional neural signal decoding algorithms often suffer from limited generalizability and high ada...
302. Accelerating Evolutionary Strategy via Rao-Blackwellizing Realization of Uncertain Input ​
Author: So Nakashima, Tetsuya J. Kobayashi
Published: 8/4/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2608.02073v1 Announce Type: cross Abstract: We investigate Optimization under Input Uncertainty (OIU), in which the input to the objective function, rather than the objective function itself, is subject to uncertainty. OIU appears in manufacturing processes with production tolerance, control o...
303. Pretraining on Call Graphs: When Binary Analysis Tasks Profit From Context ​
Author: Samuel Valenzuela, Johannes Kinder
Published: 8/4/2026, 4:00:00 AM
Categories: cs.SE, cs.CR, cs.LG
arXiv:2608.02084v1 Announce Type: cross Abstract: Binary function embedding models are trained to encode the semantics of binary code in such a way that they can be generalized to a variety of reverse engineering tasks, such as binary code search, vulnerability detection, or malware classification. ...
304. Instruction-Conditioned Exploration with Asymmetric Reinforcement Learning and Self-Distillation ​
Author: Jim Dilkes, Vahid Yazdanpanah, Sebastian Stein
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.02087v1 Announce Type: cross Abstract: Post-training Large Language Models (LLMs) with Reinforcement Learning (RL) has become an important tool for improving model capabilities, but the LLM action-space structure introduces challenges distinct from classical RL, with implications for indu...
305. From Information to Delegation: Mapping Human-AI Financial Decision Making ​
Author: Iman Munire Bilal, Yingcan Carol Wang, Ajan Raj, Filippo Giovagnini, Pranav Tewari, Yuwei Zhang, Mei-Chen Zoe Liou, Qamar Zaman
Published: 8/4/2026, 4:00:00 AM
Categories: cs.HC, cs.LG
arXiv:2608.02100v1 Announce Type: cross Abstract: As AI increasingly participates in human decision making, understanding how decision-making authority is distributed between humans and AI has become a fundamental behavioural question. We introduce a behavioural measurement framework combining inten...
306. Cardiovascular Digital Twins from Physics Based to Data Driven Approaches ​
Author: Emmanuel Lwele, Francis Chikweto
Published: 8/4/2026, 4:00:00 AM
Categories: physics.med-ph, cs.LG
arXiv:2608.02135v1 Announce Type: cross Abstract: Cardiovascular digital twins aim to create patient-specific computational models that evolve with clinical data to support diagnosis, prognosis, and therapy optimisation. Mechanistic models provide physiological interpretability but remain computatio...
307. Self-Improving Large Language Models via Progressive Experience Evolution ​
Author: Shijie Ren, Xiting Wang, Meng Li, Yujie Guo, Yunhang Yao, Ziheng Peng, Xunlong Wang, Yuetan Chen, Haoyang Zhou, Yunlong Liang, Fandong Meng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.02139v2 Announce Type: cross Abstract: Large language models (LLMs) capable of self-improvement require not only effective policy optimization, but also a principled mechanism for transforming transient interaction experience into persistent model capabilities. Existing self-improvement p...
308. CARNet: Channel-Adaptive Receiver Network for Robust NextG Communications ​
Author: Chao Jiang, Zhuo Xu, Yongli Yan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT
arXiv:2608.02172v1 Announce Type: cross Abstract: Neural receivers have been recognized as a promising paradigm for the next-generation (NextG) communications. However, due to the reliance on a static network optimized for specific channel conditions, their generalization capability across diverse s...
309. Randomized Algorithms for Learning Partitions with Near Optimal Query Complexity in Constant Rounds ​
Author: Deeparnab Chakrabarty, Aditi Dudeja, David Saulpic
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2608.02176v1 Announce Type: cross Abstract: We study the round complexity of learning a hidden partition $\mathcal{P}$ of an $n$-element universe using PAIR queries: PAIR($x,y$) tells us whether $x$ and $y$ belong to the same part of the partition or not. While it is easy to learn using $n|\ma...
310. Fast Discovery of Inclusion Dependencies with Desbordante ​
Author: Alexander Smirnov, Anton Chizhov, Ilya Shchuckin, Nikita Bobrov, George Chernishev
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DB, cs.AI, cs.DC, cs.LG, cs.PF
arXiv:2608.02213v1 Announce Type: cross Abstract: Inclusion dependency is a relation between attributes of tables that indicates possible Primary Key-Foreign Key references. Automatic discovery of inclusion dependencies is a relevant problem for both academic and industrial communities. The core con...
311. Domain-Specific Evaluation of Text-to-Speech Systems: A Multi-Metric Benchmarking Study ​
Author: Ali Jafar, Amal Sarmad, Shifa Yousaf, Maryam Bashir
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.02235v1 Announce Type: cross Abstract: Recent advances in neural text-to-speech (TTS) systems have substantially improved speech naturalness and intelligibility across many languages. However, comprehensive evaluation methodologies that jointly assess perceptual quality, speaker similarit...
312. Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability ​
Author: Abdullah Mamun, Shovito Barua Soumma, Hassan Ghasemzadeh
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.02238v1 Announce Type: cross Abstract: Ensuring trust in AI systems is essential for the safe and ethical integration of machine learning systems into high-stakes domains such as digital health. Key dimensions, including robustness, explainability, fairness, accountability, and privacy, n...
313. Self-Certification of Representation Adequacy: Sequential Certification at Minimum Task Loss ​
Author: Zijie Huang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.02267v1 Announce Type: cross Abstract: Agents that act on a compressed representation of their history face a structural risk: if the representation aliases histories with different optimal actions, no rule measurable with respect to the representation can avoid an irreducible per-round l...
314. A Multi-Objective AutoML-based Efficient Intrusion Detection System for EV Charging Networks ​
Author: Li Yang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.02274v1 Announce Type: cross Abstract: Electric Vehicle Charging Systems (EVCSs) are increasingly connected with Internet of Things (IoT) devices, which improves charging intelligence but also expands their exposure to cyber-attacks. Intrusion Detection Systems (IDSs) are essential for se...
315. Extended Field of View Analysis for VideoGAN-based Trajectory Generation ​
Author: Annajoyce Mariani, Kira Maag, Hanno Gottschalk
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.02289v1 Announce Type: cross Abstract: Realistic and diverse trajectory generation is central to enabling higher levels of vehicle automation. While rule-based and classical learning-based methods may struggle to capture the complexity of traffic behavior, generative models have already d...
316. Trajectories That Segment Themselves: Agent-Declared Boundaries as a Training Unit ​
Author: Jingxi Wei
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE
arXiv:2608.02302v1 Announce Type: cross Abstract: Long-horizon coding-agent trajectories are poorly matched to the credit units available to train on: a single action has no stable value, an episode label merges productive exploration with abandoned directions, and a fixed window cuts where the logg...
317. The Push-Forward Transform for Continuous and Robust Comparison of Dynamic Shapes ​
Author: Roua Rouatbi, Juan-Esteban Suarez Cardona, Ivo F. Sbalzarini
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.NA, math.NA
arXiv:2608.02306v1 Announce Type: cross Abstract: We introduce a mathematical framework for shape comparison based on mapping functions from the shape domain to a common reference domain. This Push-Forward Transform enables invariant and robust comparison of shapes, preserving intrinsic geometric in...
318. FastGFDs: Efficient Validation of Graph Functional Dependencies with Desbordante ​
Author: Anton Chernikov, Yurii Litvinov, Kirill Smirnov, George Chernishev
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DB, cs.AI, cs.LG, cs.PF
arXiv:2608.02321v1 Announce Type: cross Abstract: Graph functional dependencies (GFD) are a recently-developed concept aimed at capturing both topological structures in graphs and functional dependencies between attributes. The process of verifying whether a given GFD holds over a particular graph i...
319. Hard Constraints, Smooth Gradients: Learning Feasible Inventory Policies via Differentiable Projection ​
Author: Patrick Helm, Jan-Niklas Doerr, Joren Gijsbrechts, Stefan Minner
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.02343v1 Announce Type: cross Abstract: Many operational problems are constrained sequential decision processes with large, combinatorial action spaces and interdependent feasibility constraints. Mixed-integer linear programs (MILPs) handle such constraints flexibly but scale poorly in sto...
320. Self-Supervised Representations for Binary Program Clustering: From Empirical Study to Retrieval-Augmented Learning ​
Author: Martin Mocko, Daniela Chud'a
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.02348v1 Announce Type: cross Abstract: Malware clustering is a critical task in cybersecurity that helps discover threats and analyze evolving malware families. While self-supervised learning (SSL) and tabular representation learning (TRL) have achieved breakthroughs in other domains, the...
321. Faster-WAM: Do World Action Models Need Deep Action Modules? ​
Author: Liheng Ma, Rui Heng Yang, Zhanguang Zhang, Mateo Clemente, Ziwen Hu, Tongtong Cao, Yingxue Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.RO
arXiv:2608.02365v1 Announce Type: cross Abstract: World Action Models (WAMs) couple robot action prediction with video world models. Existing WAMs with shared-backbone and Mixture-of-Transformers designs generally tie the depth of the action module to that of the video backbone, resulting in substan...
322. A Spectral Filtering Approach to Regret Analysis of Distributed Online Control for Linear Dynamical Systems ​
Author: Ting-Jui Chang
Published: 8/4/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.SY, eess.SY
arXiv:2608.02375v1 Announce Type: cross Abstract: This paper studies the distributed online control problem over a network of linear time-invariant (LTI) systems in the presence of adversarial disturbances and time-varying convex costs. The network cost is characterized by the summation of local cos...
323. Cooperative Coevolution for Resource-Constrained Agentic LLM Post-Training ​
Author: Zhiyuan Wang, Shengcai Liu, Jiahao Wu, Ning Lu, Hui Ouyang, Shaofeng Zhang, Haoze Lv, Ke Tang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.02391v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents produce long, multi-turn trajectories, making gradient-based post-training memory-intensive. Evolution strategies (ES) enable memory-efficient full-parameter post-training without backpropagation and can e...
324. Network Information Enhances Unreliable News Domain Detection ​
Author: Raphaela Ke{\ss}ler, Roman David Ventzke, Viola Priesemann, Giordano De Marzo
Published: 8/4/2026, 4:00:00 AM
Categories: cs.SI, cs.LG, physics.soc-ph
arXiv:2608.02399v1 Announce Type: cross Abstract: Content-based detection of unreliable news is increasingly difficult, as low-reliability sources mimic credible journalism and generative AI makes fabricated content harder to flag. We ask whether network structure can improve news reliability classi...
325. Human-Centered Reflections on Care Robots: A Comparative Study of Caregiver Perspectives ​
Author: Laura Londo~no, Klaus Baumann, Abhinav Valada, Markus Langer
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.ET, cs.LG
arXiv:2608.02411v1 Announce Type: cross Abstract: Care robots are increasingly being introduced into healthcare settings, raising important questions about their acceptance and ethical implementation. To better understand these challenges, this study investigates caregivers' perceptions of four cate...
326. Wasserstein mixing time of the unadjusted Langevin algorithm ​
Author: Francesco Pedrotti, Peter A. Whalley
Published: 8/4/2026, 4:00:00 AM
Categories: stat.CO, cs.LG, cs.NA, math.NA, math.PR
arXiv:2608.02430v1 Announce Type: cross Abstract: We provide new estimates in Wasserstein distance for the asymptotic bias of the unadjusted Langevin algorithm, in the classical setting of log-smooth strongly log-concave measures. Our bound implies a Wasserstein mixing time of order $\kappa \sqrt{d}...
327. Intention Inference Under Execution Noise: Separating Aleatoric and Epistemic Uncertainty in Social Dilemmas ​
Author: Kival Mahadew, Jonathan Shock
Published: 8/4/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2608.02440v1 Announce Type: cross Abstract: In noisy social dilemmas, intended actions are stochastically corrupted before execution, so an observed defection may reflect hostile intent or action error. Standard Markov Decision Process (MDP) formulations treat executed actions as states, struc...
328. Advancing Relevance Measurement with Vision-Language Models for Web-Scale Search ​
Author: Han Wang, Alex Whitworth, Pak Ming Cheung, Zhenjie Zhang, Krishna Kamath, Xi Chen, Roberto Konow, Kurchi Subhra Hazra
Published: 8/4/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.02446v1 Announce Type: cross Abstract: Relevance evaluation plays a crucial role in personalized search systems, serving as a guardrail alongside user engagement metrics to ensure that search results align with user queries and intent. While human annotation is the traditional method for ...
329. Real-Time Detection and Repair of LLM Agent Failures ​
Author: Sunny Dubey
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE
arXiv:2608.02464v1 Announce Type: cross Abstract: LLM agents fail mid-episode -- they loop, cascade tool errors, drift off goal, fabricate results, or silently absorb corrupted content -- and the standard remedy, judging every step with a second LLM, costs more than the agent itself. We ask how much...
330. Private Generative Bootstrap via Blocking ​
Author: Jinwon Sohn, Veronika Ro\v{c}kov'a
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2608.02480v1 Announce Type: cross Abstract: With AI systems gaining more access to individuals' information, it is important to protect privacy when reporting statistical answers. Equally important is to privatize the reporting of uncertainty in such answers. To this end, we adopt a Bayesian l...
331. Cultural Awareness is Represented but Not Decoded: Tracing Mythological Knowledge across 18 Open-Source LLMs ​
Author: Iaroslav Chelombitko, Ekaterina Chelombitko, Mika H"am"al"ainen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.CY, cs.LG
arXiv:2608.02486v1 Announce Type: cross Abstract: Open-source LLMs reliably name Zeus, Jupiter, and Thor, but recover their counterparts in less-represented traditions like Finnish, Slavic, Egyptian, or Chinese mythology far less consistently. We ask where inside the model this cultural default is p...
332. Computational and Statistical Guarantees of the \textit{c}-Rectified flow ​
Author: Leda Wang, Zhehao Xu, Qiang Liu, Harrison H. Zhou
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC, math.PR, math.ST, stat.TH
arXiv:2608.02487v1 Announce Type: cross Abstract: Recently, rectified flow has emerged as a fundamental framework for large-scale image generation, powering state-of-the-art systems such as FLUX.1 and Stable Diffusion 3. Despite its remarkable empirical success, the computational and statistical gua...
333. Beyond Modern Asymptotics for Log-Likelihood Ratios in Logistic Regression ​
Author: Hugo Chardon, Reese Pathak, Nikita Zhivotovskiy
Published: 8/4/2026, 4:00:00 AM
Categories: math.ST, cs.IT, cs.LG, math.IT, stat.TH
arXiv:2608.02507v1 Announce Type: cross Abstract: We characterize the finite sample behavior of the log-likelihood ratio statistic in binary logistic regression, uniformly over both the design and the target parameter. For $n\geq d\geq 3$, we determine, up to universal constants, its worst case $(1-...
334. Optimizing Minimax Regret in Uncertain MDPs with Small Sets of Policies ​
Author: Sterre Lutz, Dani"el Vos, Matthijs T. J. Spaan, Anna Lukina
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.02509v1 Announce Type: cross Abstract: Sequential decision-making in real-world applications often involves uncertainty about the environment's model. Uncertain Markov decision processes (UMDPs) represent the possible environments as a set of MDPs with shared states and actions but potent...
335. LiveMem: Maintaining Memory State Continuity in Long-Running LLM Inference ​
Author: Zhichen Liu, Ruihan Sun, Hengjie Yang, Zipeng Wu, Zhaohan Chen, Xiaofan Zhang, Yang Xu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.02515v1 Announce Type: cross Abstract: Long-running assistants and agents consume interaction streams that eventually outgrow the context. Existing context retention, summarization, and retrieval preserve access to selected history, but do not provide a persistent state over the full life...
336. Optimal Unambiguous DNFs and Alon-Saks-Seymour ​
Author: Chirag Pabbaraju
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CC, cs.DM, cs.LG
arXiv:2608.02533v1 Announce Type: cross Abstract: We construct unambiguous DNFs having width $O(n)$ but $0$-certificate complexity $\Omega(n^2)$. By utilizing the special structure of these DNFs, we prove a lifting theorem with a constant-sized gadget that lifts the DNF to a communication problem, w...
337. Interaction Is Not Necessary for Order-Optimal 1-Bit Mean Estimation ​
Author: Jiachen Hu, Han Zhong
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT, math.ST, stat.TH
arXiv:2608.02538v1 Announce Type: cross Abstract: This paper is concerned with one-bit mean estimation, where each independent sample is represented by a single binary message. We consider distributions on $\mathbb{R}$ with mean in $[-\lambda,\lambda]$ and absolute $k$-th central moment at most $\si...
338. A Simple Approximation to the Distribution of the Ridge Regression Estimator ​
Author: Jos'e Luis Montiel Olea, Ryan Strong, Amilcar Velez, Zhuoheng Xu, Haomin Yu
Published: 8/4/2026, 4:00:00 AM
Categories: econ.EM, cs.LG, math.ST, stat.TH
arXiv:2608.02539v1 Announce Type: cross Abstract: We present a simple Gaussian approximation to the finite-sample distribution of the classical ridge regression estimator. Our approximation captures the fact that, in finite samples, the ridge regression estimator trades off bias and variance to redu...
339. CoWAM: Coordination Contracts for Selective Policy Intervention with WAMs ​
Author: Shuaijun Liu, Qifu Wen, Shuyang Hao, Qi Luo, Chenglong Zhang, Feiyang You, Chengyu Wu, Ningxin Su
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2608.02578v1 Announce Type: cross Abstract: World Action Models (WAMs) augment robot policies with action-conditioned predicted futures, but a plausible future alone does not justify changing the action that a bimanual policy would execute. We present CoWAM, a selective intervention layer that...
340. The Condition-Number Barrier in Sparse Least Squares ​
Author: Honghao Lin, Vahab Mirrokni, David P. Woodruff
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2608.02588v1 Announce Type: cross Abstract: In [AS21], Axiotis and Sviridenko conjectured that the linear dependence on the restricted condition number in sparse convex optimization cannot be improved by a polynomial-time algorithm. We establish their conjectured lower bound for least-squares ...
341. The Elements of Differentiable Programming ​
Author: Mathieu Blondel, Vincent Roulet
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.PL
arXiv:2403.14606v4 Announce Type: replace Abstract: Artificial intelligence has recently experienced remarkable advances, fueled by large models, vast datasets, accelerated hardware, and, last but not least, the transformative power of differentiable programming. This new programming paradigm enable...
342. Neural Surrogate HMC: On Using Neural Likelihoods for Hamiltonian Monte Carlo in Simulation-Based Inference ​
Author: Linnea M Wolniewicz, Peter Sadowski, Claudio Corti
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.HE
arXiv:2407.20432v3 Announce Type: replace Abstract: Bayesian inference methods such as Markov Chain Monte Carlo (MCMC) typically require repeated computations of the likelihood function, but in some scenarios this is infeasible and alternative methods are needed. Simulation-based inference (SBI) met...
343. Efficient nonlinear flame response modeling for propulsion thermoacoustic analysis using limited numerical data ​
Author: Jiawei Wu, Teng Wang, Jiaqi Nan, Wang Han, Lijun Yang, Jingxuan Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2409.05885v2 Announce Type: replace Abstract: Characterizing nonlinear flame response is critical for predicting thermoacoustic instabilities in propulsion combustors, yet obtaining a comprehensive response map through high-fidelity simulations remains computationally prohibitive. This study p...
344. Polyatomic Complexes: A topologically-informed learning representation for atomistic systems ​
Author: Rahul Khorana, Marcus Noack, Jin Qian
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2409.15600v3 Announce Type: replace Abstract: A representation of a molecule or material should be invariant to the symmetries of physics, unique, continuous, efficient and general. These properties, however, are hard to satisfy at once: a descriptor invariant under the full orthogonal group $...
345. Sparse Covariance Neural Networks ​
Author: Andrea Cavallo, Zhan Gao, Elvin Isufi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2410.01669v3 Announce Type: replace Abstract: Covariance Neural Networks (VNNs) perform graph convolutions on the covariance matrix of input data to leverage correlation information as pairwise connections. They have achieved success in a multitude of applications such as neuroscience, financi...
346. Defending Membership Inference Attacks via Privacy-aware Sparsity Tuning ​
Author: Hengxiang Zhang, Qiang Hu, Hongxin Wei
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2410.06814v2 Announce Type: replace Abstract: Over-parameterized models are typically vulnerable to membership inference attacks, which aim to determine whether a specific sample is included in the training of a given model. Previous Weight regularizations (e.g., L1 regularization) typically i...
347. Belief-Contraction-Driven Active Inverse Source Localization and Characterization ​
Author: Yiwei Shi, Mengyue Yang, Qi Zhang, Cunjia Liu, Weinan Zhang, Weiru Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2501.13084v2 Announce Type: replace Abstract: Active inverse source localization and characterization (ISLC) in dynamic fields requires sequential decision making under partial observability, where a mobile sensor must infer latent source parameters from sparse, noisy readings. We introduce a ...
348. Development and Validation of a Dynamic Kidney Failure Prediction Model based on Deep Learning: A Real-World Study with External Validation ​
Author: Jingying Ma, Jinwei Wang, Lanlan Lu, Zhiqin Jiang, Mengling Feng, Feifei Zhang, Peng Shen, Yexiang Sun, Shenda Hong, Luxia Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.AP
arXiv:2501.16388v3 Announce Type: replace Abstract: Background: Chronic kidney disease (CKD), a progressive disease with high morbidity and mortality, has become a significant global public health problem. Most existing models are static and fail to capture temporal trends in disease progression, li...
349. GradientStabilizer:Fix the Norm, Not the Gradient ​
Author: Tianjin Huang, Zhangyang Wang, Haotian Hu, Zhenyu Zhang, Gaojie Jin, Xiang Li, Li Shen, Jiaxing Shang, Tianlong Chen, Ke Li, Lu Liu, Qingsong Wen, Shiwei Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2502.17055v5 Announce Type: replace Abstract: Training instability in modern deep learning systems is frequently triggered by rare but extreme gradient-norm spikes, which can induce oversized parameter updates, corrupt optimizer state, and lead to slow recovery or divergence. Widely used safeg...
350. MUSS: Multilevel Subset Selection for Relevance and Diversity ​
Author: Vu Nguyen, Andrey Kan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2503.11126v4 Announce Type: replace Abstract: The problem of relevant and diverse subset selection has a wide range of applications, including recommender systems and retrieval-augmented generation (RAG). For example, in recommender systems, one is interested in selecting relevant items, while...
351. Sampling Decisions: Exact Path-Space Control for Physics-Informed Generative Sampling ​
Author: Michael Chertkov, Hamidreza Behjoo, Sungsoo Ahn
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, cs.AI, cs.SY, eess.SY, stat.ML
arXiv:2503.14549v4 Announce Type: replace Abstract: Scientific generative models must turn tractable local decisions into globally correlated samples that respect physical constraints. We introduce Sampling Decisions, a finite-horizon framework in which a structured object is assembled on a growing ...
352. Understanding Machine Unlearning Through the Lens of Mode Connectivity ​
Author: Jiali Cheng, Hadi Amiri
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CV
arXiv:2504.06407v2 Announce Type: replace Abstract: Machine Unlearning aims to remove undesired information from trained models without full retraining from scratch. Despite recent progress, the loss landscape and optimization geometry of unlearning are poorly understood. In this paper, we study mac...
353. Fairness in Augmented Graph Learning: A Survey ​
Author: Renqiang Luo, Huafei Huang, Ziqi Xu, Xikun Zhang, Enyan Dai, Bo Yang, Feng Xia
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2504.21296v2 Announce Type: replace Abstract: Graph learning has evolved into Augmented Graph Learning (AGL) by integrating specialized machine learning (ML) techniques. Examples include federated learning, graph transformers, and graph condensation. While enhancing model utility, AGL introduc...
354. Credit Assignment and Efficient Exploration based on Influence Scope in Multi-agent Reinforcement Learning ​
Author: Shuai Han, Mehdi Dastani, Shihan Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2505.08630v2 Announce Type: replace Abstract: Training cooperative agents in sparse-reward scenarios poses significant challenges for multi-agent reinforcement learning (MARL). Without clear feedback on actions at each step in sparse-reward setting, previous methods struggle with precise credi...
355. Training Deep Morphological Neural Networks as Universal Approximators ​
Author: Konstantinos Fotopoulos, Petros Maragos
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2505.09710v4 Announce Type: replace Abstract: We investigate deep morphological neural networks (DMNNs), studying how changes in algebraic structure affect the expressivity and trainability of deep architectures. We show that despite the inherent non-linearity of morphological operations, exis...
356. DeepConvContext: A Multi-Scale Approach to Timeseries Classification in Human Activity Recognition ​
Author: Marius Bock, Juergen Gall, Michael Moeller, Kristof Van Laerhoven
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, eess.IV
arXiv:2505.20894v2 Announce Type: replace Abstract: Despite recognized limitations in modeling long-range temporal dependencies, Human Activity Recognition (HAR) has traditionally relied on a sliding window approach to segment labeled datasets. Deep learning models like the DeepConvLSTM typically cl...
357. EHR2Path: Comprehensive Pathway-Level Modeling of Longitudinal Patient Trajectories from Multimodal Electronic Health Records ​
Author: Chantal Pellegrini, Ege "Ozsoy, David Bani-Harouni, Matthias Keicher, Nassir Navab
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2506.04831v3 Announce Type: replace Abstract: Forecasting how a patient's condition is likely to evolve, including possible deterioration, recovery, treatment needs, and care transitions, could support more proactive and personalized care, but requires modeling heterogeneous and longitudinal e...
358. Quality-Diversity Red-Teaming: Automated Generation of High-Quality and Diverse Attackers for Large Language Models ​
Author: Ren-Jian Wang, Ke Xue, Zeyu Qin, Ziniu Li, Sheng Tang, Hao-Tian Li, Shengcai Liu, Zhi Yu, Yuanpeng Tan, Chao Qian
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2506.07121v2 Announce Type: replace Abstract: Ensuring the safety and robustness of large language models (LLMs) is a fundamental challenge and a critical prerequisite for the responsible deployment of artificial intelligence. Red-teaming, a systematic framework to identify adversarial prompts...
359. ARMOR: Robust Reinforcement Learning-based Control for UAVs under Physical Attacks ​
Author: Pritam Dash, Ethan Chan, Nathan P. Lawrence, Karthik Pattabiraman
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.RO
arXiv:2506.22423v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles (UAVs) depend on onboard sensors for perception, navigation, and control. However, these sensors are susceptible to physical attacks, such as GPS spoofing, that can corrupt state estimates and lead to unsafe behavior. While...
360. Learning Graph-Indexed Trajectory Patterns for Stochastic On-Time Arrival Routing ​
Author: Yuanhang Wang, Xing Wei, Duoxiang Zhao, Zezhou Zhang, Hao Qin, Yuqi Ouyang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2508.17218v4 Announce Type: replace Abstract: Correlated link travel times create decision-relevant patterns in partial route histories. In stochastic on-time arrival (SOTA) routing, each route prefix forms a variable-length, graph-indexed sequence in which traversed-edge identities, realized ...
361. Self-composing neural operators for high-frequency and multiscale PDE surrogates ​
Author: Juncai He, Xinliang Liu, Jinchao Xu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2508.20650v2 Announce Type: replace Abstract: Addressing the computational challenges of high-frequency and multiscale partial differential equations (PDEs), this work introduces a self-composing neural operator (SC-NO) framework. Inspired by classical fixed-point iterative solvers (e.g., mult...
362. CAPMix: Robust KPI Anomaly Detection for AIOps in Noisy and Dynamic Environments ​
Author: Xudong Mou, Rui Wang, Tiejun Wang, Zexin Wu, Fangda Guo, Jie Sun, Shiru Chen, Penghao Zhang, Tiezi Zhang, Tianyu Wo, Hao Peng, Chunming Hu, Xudong Liu, Renyu Yang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2509.06419v2 Announce Type: replace Abstract: Time-series anomaly detection is crucial in AIOps for maintaining large-scale service reliability. In production, streams of Key Performance Indicators (KPI) are high-dimensional, non-stationary, and affected by noise, deployment changes, and laten...
363. Breaking the Statistical Similarity Trap in Extreme Convection Detection ​
Author: Md Tanveer Hossain Munim
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2509.09195v2 Announce Type: replace Abstract: Current evaluation metrics for deep learning weather models create a "Statistical Similarity Trap", rewarding blurry predictions while missing rare, high-impact events. We provide quantitative evidence of this trap, showing sophisticated baselines ...
364. CountTRuCoLa: Rule Learning for Interpretable Temporal Knowledge Graph Forecasting ​
Author: Julia Gastinger, Christian Meilicke, Heiner Stuckenschmidt
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.09474v2 Announce Type: replace Abstract: We address the task of temporal knowledge graph forecasting with an inherently interpretable method based on symbolic rules. Motivated by recent work proposing a strong baseline based on recurrent facts, our approach learns four simple rule types, ...
365. Learned Digital Over-the-Air Computing for Federated Edge Learning ​
Author: Antonio Tarizzo, Mohammad Kazemi, Deniz G"und"uz
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2509.16577v2 Announce Type: replace Abstract: Over-the-air (OTA) aggregation enables federated edge learning (FEEL) by exploiting the superposition property of the wireless channel to merge communication with computation, eliminating the need to schedule and decode devices individually. Analog...
366. SingLEM: Single-Channel Large EEG Model ​
Author: Jamiyan Sukhbaatar, Satoshi Imamura, Ibuki Inoue, Shoya Murakami, Kazi Mahmudul Hassan, Seungwoo Han, Ingon Chanpornpakdi, Toshihisa Tanaka
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.17920v2 Announce Type: replace Abstract: Current deep learning models for electroencephalography (EEG) are often task-specific and depend on large labeled datasets, limiting their adaptability. Although EEG foundation models seek broader applicability, many still rely on predefined multi-...
367. T-TAMER: Provably Taming Trade-offs in ML Serving ​
Author: Yuanyuan Yang, Ruimin Zhang, Jamie Morgenstern, Haifeng Xu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.GT
arXiv:2509.22992v2 Announce Type: replace Abstract: As machine learning models continue to grow in size and complexity, efficient serving faces increasingly broad trade-offs spanning accuracy, latency, resource usage, and other objectives. Multi-model serving further complicates these trade-offs; fo...
368. Refine Drugs, Don't Complete Them: Uniform-Source Discrete Flows for Fragment-Based Drug Discovery ​
Author: Benno Kaech, Luis Wyss, Karsten Borgwardt, Gianvito Grasso
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.26405v2 Announce Type: replace Abstract: We introduce InVirtuoGen, a discrete flow generative model for fragmented SMILES for de novo and fragment-constrained generation, and target-property/lead optimization of small molecules. The model learns to transform a uniform source over all poss...
369. Surrogate Modeling for the Design of Optimal Lattice Structures using Tensor Completion ​
Author: Shaan Pakala, Aldair E. Gongora, Brian Giera, Evangelos E. Papalexakis
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.07474v2 Announce Type: replace Abstract: When designing new materials, it is often necessary to design a material with specific desired properties. Unfortunately, as new design variables are added, the search space grows exponentially, which makes synthesizing and validating the propertie...
370. Eigenvalues as a Metric for Memory Dynamics in Sequence Models ​
Author: Rahel Rickenbach, Jelena Trisovic, Alexandre Didier, Jerome Sieber, Melanie N. Zeilinger
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY
arXiv:2510.09379v2 Announce Type: replace Abstract: While softmax attention drives state-of-the-art performance in sequence modeling, its quadratic complexity motivates linear alternatives such as state space models (SSMs). Structural differences between the two model classes, however, hinder direct...
371. NashPG: A Policy Gradient Method with Iteratively Refined Regularization for Finding Nash Equilibria ​
Author: Eason Yu, Tzu Hao Liu, Cl'ement L. Canonne, Yunke Wang, Chang Xu, Nguyen H. Tran, Stefano V. Albrecht
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.GT
arXiv:2510.18183v3 Announce Type: replace Abstract: Finding Nash equilibria in two-player zero-sum imperfect-information games remains a central challenge in multi-agent reinforcement learning. Recent multi-round regularization methods offer a promising direction, yet existing approaches either requ...
372. Estimating Treatment Effects in Networks under Unknown Exposure Mappings ​
Author: Daan Caljon, Jente Van Belle, Wouter Verbeke
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.21457v2 Announce Type: replace Abstract: Estimating heterogeneous treatment effects in network settings is complicated by interference, meaning that the outcome of an instance can be influenced by the treatment status of others. Existing causal machine learning approaches that account for...
373. LLM generation novelty through the lens of semantic similarity ​
Author: Philipp Davydov, Ameya Prabhu, Matthias Bethge, Elisa Nguyen, Seong Joon Oh
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2510.27313v3 Announce Type: replace Abstract: Generation novelty is a key indicator of an LLM's ability to generalize, yet measuring it against full pretraining corpora is computationally challenging. Existing evaluations often rely on lexical overlap, failing to detect paraphrased text, or do...
374. Extending Fair Null-Space Projections for Continuous Attributes to Kernel Methods ​
Author: Felix St"orck, Fabian Hinder, Barbara Hammer
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2511.03304v3 Announce Type: replace Abstract: With the on-going integration of machine learning systems into the everyday social life of millions the notion of fairness becomes an ever increasing priority in their development. Fairness notions commonly rely on protected attributes to assess po...
375. Regularized Schr\"odinger Bridge via Distortion-Perception Perturbation for High-Fidelity Speech Enhancement ​
Author: Qing Yao, Lijian Gao, Qirong Mao, Ming Dong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.SD
arXiv:2511.11686v4 Announce Type: replace Abstract: Speech enhancement (SE) requires high-fidelity reconstruction of clean speech that preserves linguistic and paralinguistic cues while maintaining high perceptual quality. Recently, Schr"odinger Bridge (SB), a family of diffusion-based generative m...
376. LAYA: Layer-wise Attention Aggregation for Interpretable Depth-Aware Neural Networks ​
Author: Gennaro Vessio
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.12723v2 Announce Type: replace Abstract: Deep neural networks typically rely on the representation produced by their final hidden layer to make predictions, implicitly assuming that this single vector fully captures the semantics encoded across all preceding transformations. However, inte...
377. DeepDefense: Robust Learning via Layer-Wise Gradient-Feature Alignment ​
Author: Ci Lin, Tet Yeap, Iluju Kiringa
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2511.13749v2 Announce Type: replace Abstract: Deep neural networks are known to be vulnerable to adversarial perturbations, which are small, carefully crafted inputs that lead to incorrect predictions. In this paper, we propose DeepDefense, a novel defense framework that applies Gradient-Featu...
378. Amortized Inference of Multi-Modal Posteriors using Likelihood-Weighted Normalizing Flows ​
Author: Rajneil Baruah
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, hep-ex, hep-ph, physics.comp-ph, physics.data-an
arXiv:2512.04954v3 Announce Type: replace Abstract: We present a novel technique for amortized posterior estimation using Normalizing Flows trained with likelihood-weighted importance sampling. This approach allows for the efficient inference of theoretical parameters in high-dimensional inverse pro...
379. Auto-exploration for online reinforcement learning ​
Author: Caleb Ju, Guanghui Lan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC
arXiv:2512.06244v3 Announce Type: replace Abstract: The exploration-exploitation dilemma in reinforcement learning (RL) is a fundamental challenge to efficient RL algorithms. Existing algorithms for finite state and action discounted RL problems address this by assuming sufficient exploration over b...
380. Conformal bandits: bringing statistical validity and reward efficiency under weak arm separability ​
Author: Simone Cuonzo, Nina Deliu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.09850v2 Announce Type: replace Abstract: We introduce Conformal Bandits, a novel framework integrating Conformal Prediction (CP) into bandit problems, a classic paradigm for sequential decision-making under uncertainty. Traditional regret-minimisation bandit strategies like Thompson Sampl...
381. EEG-FM-Compass: Progress, Benchmarking, and Future Directions for EEG Foundation Models ​
Author: Dingkun Liu, Yuheng Chen, Zhu Chen, Zhenyao Cui, Yaozhi Wen, Jiayu An, Jingwei Luo, Dongrui Wu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2601.17883v3 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models (FMs) have recently emerged as a promising paradigm for brain-computer interfaces, aiming to learn transferable neural representations from large-scale heterogeneous recordings. Despite rapid progress,...
382. Rethinking Federated Graph Foundation Models: A Graph-Language Alignment-based Approach ​
Author: Yinlin Zhu, Di Wu, Xianzhi Zhang, Yuming Ai, Xunkai Li, Miao Hu, Guocong Quan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.21369v2 Announce Type: replace Abstract: Recent studies of federated graph foundational models (FedGFMs) break the idealized and untenable assumption of having centralized data storage to train graph foundation models, and accommodate the reality of distributed, privacy-restricted data si...
383. Understanding Rate-Distortion Performance in Distributed Transformer Inference ​
Author: Anderson de Andrade, Alon Harell, Ivan V. Baji'c
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2601.22002v5 Announce Type: replace Abstract: Transformers achieve superior performance on many tasks, but impose heavy compute and memory requirements during inference. This inference can be made more efficient by partitioning the process across multiple devices, which, in turn, requires comp...
384. AROpt: An Optimization Method for Autoregressive Time Series Forecasting ​
Author: Zheng Li, Jerry Cheng, Huanying Gu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.02288v3 Announce Type: replace Abstract: Current time-series forecasting models are primarily based on transformer-style neural networks. These models achieve long-term forecasting mainly by scaling up the model size rather than through genuinely autoregressive (AR) rollout. From the pers...
385. RAP: KV-Cache Compression via RoPE-Aligned Pruning ​
Author: Jihao Xin, Tian Lyu, David Keyes, Hatem Ltaief, Marco Canini
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.02599v4 Announce Type: replace Abstract: Long-context inference in large language models (LLMs) is bottlenecked by the memory and compute of the key-value (KV) cache. Structured pruning is a direct way to shrink it: dropping the least useful channels of the W_k, W_v projection weights to ...
386. Superposition Without Interference? Towards Isolated Interventions via Almost Orthogonal Features in Language Models ​
Author: Moritz Miller, Florent Draye, Bernhard Sch"olkopf
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2602.04718v4 Announce Type: replace Abstract: A central premise in mechanistic interpretability is that meaningful concepts in language models are represented by linear features in activation space. For such features to support reliable interventions, manipulating one feature should not substa...
387. Optimized Piecewise Affine Abstractions of Neural Networks with Learnable Activation Functions ​
Author: Noah Schwartz, Chandra Kanth Nagesh, Sriram Sankaranarayanan, Ramneet Kaur, Tuhin Sahai, Susmit Jha
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.LO
arXiv:2602.06737v2 Announce Type: replace Abstract: We present a generalized framework for the range verification of neural networks featuring non-linear activation functions. Our approach first constructs an ``optimized piecewise affine abstraction" of the network that replaces each non-linear acti...
388. Hyperparameter Transfer Laws for Non-Recurrent Multi-Path Neural Networks ​
Author: Shenxi Wu, Haosong Zhang, Xingjian Ma, Shirui Bian, Yichi Zhang, Xi Chen, Wei Lin
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.07494v2 Announce Type: replace Abstract: Deeper modern architectures are costly to train, making hyperparameter transfer preferable to expensive repeated tuning. Maximal Update Parametrization ($\mu$P) helps explain why many hyperparameters transfer across width. Yet depth scaling is less...
389. Self-Evolving Recommendation System: End-To-End Autonomous Model Optimization With LLM Agents ​
Author: Haochen Wang, Yi Wu, Daryl Chang, Li Wei, Lukasz Heldt
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.10226v3 Announce Type: replace Abstract: Optimizing large-scale machine learning systems, such as recommendation models for global video platforms, requires navigating a massive hyperparameter search space and, more critically, designing sophisticated optimizers, architectures, and reward...
390. LakeMLB: Data Lake Machine Learning Benchmark ​
Author: Feiyu Pan, Tianbin Zhang, Aoqian Zhang, Yu Sun, Zheng Wang, Lixing Chen, Li Pan, Jianhua Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.10441v2 Announce Type: replace Abstract: Data lakes have become a fundamental platform for large-scale machine learning by enabling flexible management of heterogeneous data. Despite their growing importance, standardized benchmarks for evaluating machine learning performance in data lake...
391. Token-Efficient Change Detection in LLM APIs ​
Author: Timoth'ee Chauvin, Cl'ement Lalanne, Erwan Le Merrer, Jean-Michel Loubes, Fran\c{c}ois Ta"iani, Gilles Tredan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2602.11083v4 Announce Type: replace Abstract: Remote change detection in LLMs is a difficult problem. Existing methods are either too expensive for deployment at scale, or require initial white-box access to model weights or grey-box access to log probabilities. We aim to achieve both low cost...
392. Just on Time: Token-Level Early Stopping for Diffusion Language Models ​
Author: Zakhar Kohut, Severyn Shykula, Mykola Vysotskyi, Serhii Dmytryshyn, Dmytro Khamula, Michal Zakrzewski, Damian Rynczak, Jacek Ma{\l}ecki, Taras Rumezhak, Volodymyr Karpiv
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2602.11133v2 Announce Type: replace Abstract: Diffusion language models generate text through iterative refinement, a process that is often computationally inefficient because many tokens reach stability long before the final denoising step. We introduce a training-free, token-level early stop...
393. Multi-Level Strategic Classification: Incentivizing Improvement through Promotion and Relegation Dynamics ​
Author: Ziyuan Huang, Lina Alkarmi, Mingyan Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.11439v3 Announce Type: replace Abstract: Strategic classification studies the problem where self-interested individuals or agents manipulate their response to obtain favorable decision outcomes made by classifiers, typically turning to dishonest actions when they are less costly than genu...
394. TempoNet: Slack-Quantized Transformer-Guided Reinforcement Scheduler for Adaptive Deadline-Centric Real-Time Dispatchs ​
Author: Rong Fu, Yibo Meng, Zeyu Zhang, Ziming Guo, Jia Yee Tan, Xiaojing Du, Simon James Fong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.OS, cs.SY, eess.SY
arXiv:2602.18109v4 Announce Type: replace Abstract: Real-time schedulers must reason about tight deadlines under strict compute budgets. We present TempoNet, a reinforcement learning scheduler that pairs a permutation-invariant Transformer with a deep Q-approximation. An Urgency Tokenizer discretize...
395. Large Causal Models for Temporal Causal Discovery ​
Author: Nikolaos Kougioulis, Nikolaos Gkorgkolis, MingXue Wang, Bora Caglayan, Dario Simionato, Andrea Tonon, Ioannis Tsamardinos
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.18662v2 Announce Type: replace Abstract: Causal discovery for both cross-sectional and temporal data has traditionally followed a dataset-specific paradigm, where a new model is fitted for each individual dataset. Such an approach limits the potential of multi-dataset pretraining. The con...
396. Provably Safe Generative Sampling with Constricting Barrier Functions ​
Author: Darshan Gadginmath, Ahmed Allibhoy, Fabio Pasqualetti
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY, math.OC
arXiv:2602.21429v3 Announce Type: replace Abstract: Flow-based generative models, such as diffusion models and flow matching models, have achieved remarkable success in learning complex data distributions. However, a critical gap remains for their deployment in safety-critical domains: the lack of f...
397. Local Shapley: Model-Induced Locality and Optimal Reuse in Data Valuation ​
Author: Xuan Yang, Hsi-Wen Chen, Ming-Syan Chen, Jian Pei
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DB, cs.GT
arXiv:2603.03672v2 Announce Type: replace Abstract: The Shapley value provides a principled foundation for data valuation, but exact computation is #P-hard due to the exponential coalition space. Existing accelerations remain global and ignore a structural property of modern predictors: for a given ...
398. What Makes Position Zero Special? A Mechanistic Study of Position Zero Attention Sinks in LLMs ​
Author: Runyu Peng, Ruixiao Li, Mingshu Chen, Yunhua Zhou, Qipeng Guo, Xipeng Qiu, Yucheng Lu, Chen Zhao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2603.06591v2 Announce Type: replace Abstract: Transformers frequently allocate disproportionate attention to specific tokens, a phenomenon known as attention sinks. Causal large language models reliably form one at position zero, though its role remains debated. We approach this question from ...
399. Heterogeneous Decentralized Diffusion Models ​
Author: Zhiying Jiang, Raihan Seraj, Marcos Villagra, Bidhan Roy
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2603.06741v3 Announce Type: replace Abstract: Training frontier-scale diffusion models often requires substantial computational resources concentrated in tightly-coupled clusters, limiting participation to well-resourced institutions. While Decentralized Diffusion Models (DDM) enable training ...
400. GPrune-LLM: Generalization-Aware Structured Pruning for Large Language Models ​
Author: Xiaoyun Liu, Divya Saxena, Jiannong Cao, Yuqing Zhao, Yiying Dong, Penghui Ruan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.13418v2 Announce Type: replace Abstract: Structured pruning is widely applied to compress large language models (LLMs), but its performance depends heavily on how neuron importance is estimated. Most existing methods rely on activation statistics from a single calibration set, which intro...
401. GAPSL: A Gradient-Aligned Parallel Split Learning over Data-Heterogeneous Edge Computing Systems ​
Author: Zheng Lin, Ons Aouedi, Zihan Fang, Wei Ni, Yue Gao, Symeon Chatzinotas, Xianhao Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.18540v2 Announce Type: replace Abstract: The increasing complexity of neural networks poses significant challenges for democratizing federated learning (FL) on resource-constrained edge devices. Parallel split learning (PSL) has emerged as a promising solution by offloading substantial co...
402. When Differential Privacy Meets Wireless Federated Learning: An Improved Analysis for Privacy and Convergence ​
Author: Chen Yaoling, Liang Hao, Tu Xiaotong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.19040v2 Announce Type: replace Abstract: Differentially private wireless federated learning (DPWFL) is a promising framework for protecting sensitive user data. However, foundational questions on how to precisely characterize privacy loss remain open, and existing work is further limited ...
403. GraphER: An Efficient Graph-Based Enrichment and Reranking Method for Retrieval-Augmented Generation ​
Author: Ruizhong Miao, Yuying Wang, Rongguang Wang, Chenyang Li, Tao Sheng, Sujith Ravi, Dan Roth
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IR
arXiv:2603.24925v3 Announce Type: replace Abstract: Semantic search in retrieval-augmented generation (RAG) systems is often insufficient for complex information needs, particularly when relevant evidence is scattered across multiple sources, because it may fail to retrieve the complete set of evide...
404. SIGMA: Semantic Identifier Grouping for Molecular Autoregression ​
Author: Xinyu Wang, Fei Dou, Jinbo Bi, Minghu Song
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.25062v2 Announce Type: replace Abstract: Autoregressive molecular models assign probability to molecular serializations even though chemical identity is invariant to serialization. Equivalent serializations can therefore represent a common molecular identity yet induce inconsistent next-t...
405. ARMOR: A Robust Self-Supervised Framework for Root Cause Analysis in Microservices under Missing Modality ​
Author: Wenzhuo Qian, Hailiang Zhao, Ziqi Wang, Zhipeng Gao, Jiayi Chen, Zhiwei Ling, Shuiguang Deng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.SE
arXiv:2603.25538v3 Announce Type: replace Abstract: Automated incident management is critical for microservice reliability. While recent unified frameworks leverage multimodal data for joint optimization, they unrealistically assume perfect data completeness. In practice, network fluctuations and ag...
406. From Vessel Trajectories to Safety-Critical Encounter Scenarios: A Generative AI Framework for Autonomous Ship Digital Testing ​
Author: Sijin Sun, Liangbin Zhao, Xiuju Fu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.28067v2 Announce Type: replace Abstract: Digital testing has emerged as a key paradigm for the development and verification of autonomous maritime navigation systems, yet the availability of realistic and diverse safety-critical encounter scenarios remains limited. Existing approaches eit...
407. A Perturbation Approach to Unconstrained Linear Bandits ​
Author: Andrew Jacobsen, Dorian Baudry, Shinji Ito, Nicol`o Cesa-Bianchi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2603.28201v3 Announce Type: replace Abstract: We revisit the standard perturbation-based approach of Abernethy et al. (2008) in the context of unconstrained Bandit Linear Optimization (uBLO). We show the surprising result that in the unconstrained setting, this approach effectively reduces Ban...
408. Expert-Choice Routing Enables Adaptive Computation in Diffusion Language Models ​
Author: Shuibai Zhang, Caspian Zhuang, Chihan Cui, Zhihan Yang, Fred Zhangzhi Peng, Yanxin Zhang, Haoyue Bai, Zack Jia, Yang Zhou, Guanhua Chen, Ming Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2604.01622v2 Announce Type: replace Abstract: Diffusion language models (DLMs) enable parallel, non-autoregressive text generation, yet existing DLM mixture-of-experts (MoE) models inherit token-choice (TC) routing from autoregressive systems, leading to load imbalance and rigid computation al...
409. K-STEMIT: Knowledge-Informed Spatio-Temporal Efficient Multi-Branch Graph Neural Network for Subsurface Stratigraphy Thickness Estimation from Radar Data ​
Author: Zesheng Liu, Maryam Rahnemoonfar
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2604.09922v2 Announce Type: replace Abstract: Subsurface stratigraphy contains important spatio-temporal information about accumulation, deformation, and layer formation in polar ice sheets. In particular, variations in internal ice layer thickness provide valuable constraints for snow mass ba...
410. (How) Learning Rates Regulate Catastrophic Overtraining ​
Author: Mark Rofin, Aditya Varre, Nicolas Flammarion
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2604.13627v2 Announce Type: replace Abstract: Supervised fine-tuning (SFT) is a common first stage of LLM post-training, teaching the model to follow instructions and shaping its behavior as a helpful assistant. At the same time, SFT may harm the fundamental capabilities of an LLM, particularl...
411. Do Prompt-Elicited Trajectories Reflect Training-Time Reward Hacking? A Systematic Study on Monitoring Training-Time Reward Hacking in Code Generation ​
Author: Lichen Li, Hengguang Zhou, Yijun Liang, Tianyi Zhou, Cho-Jui Hsieh
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.23488v3 Announce Type: replace Abstract: Reward hacking in code generation, where models exploit evaluation loopholes to obtain high reward without correctly solving the intended task, poses a critical challenge for Reinforcement Learning (RL) and the deployment of reasoning models. Exist...
412. Meritocratic Fairness via $K$-Shapley Values in Budgeted Combinatorial Bandits with Full-Bandit Feedback ​
Author: Shradha Sharma, Shweta Jain, Swapnil Dhamal
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA
arXiv:2605.00762v2 Announce Type: replace Abstract: We study meritocratic fairness in budgeted combinatorial multi-armed bandits with full-bandit feedback, where a learner selects at most $K$ arms per time step and observes only the noisy aggregate reward of the selected set. To define merit under b...
413. Intersectional Disentangling of Temporal and Acquisition Bias in Fetal Ultrasound ​
Author: Aya Elgebaly, Joris Fournel, Benjamin Laine J{\o}nch Jurgensen, Kamil Mikolaj, Anders Christensen, Martin Tolsgaard, Claes Ladefoged, Aasa Feragen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, eess.IV
arXiv:2605.02942v2 Announce Type: replace Abstract: Fairness studies of medical imaging AI often explain subgroup performance gaps through under-representation in the training data. We show that intersectional analysis can disentangle fairness and performance gaps arising from clinical and acquisiti...
414. Retrieval of Coastal Biogeochemical Parameters From Near-Surface Hyperspectral Remote Sensing Reflectance Using Physics-Aware Meta-Learning ​
Author: Yiqing Guo, Nagur R. C. Cherukuru, Eric A. Lehmann, S. L. Kesav Unnithan, Tim J. Malthus, Gemma Kerrisk, Xiubin Qi, Faisal Islam, Tisham Dhar, Mark J. Doubell
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.05623v3 Announce Type: replace Abstract: Hyperspectral in situ sensing has shown promise in retrieving aquatic biogeochemical (BGC) parameters, such as total suspended solids, dissolved organic carbon, and total chlorophyll-a, for cost-effective monitoring of coastal water quality. Howeve...
415. Quantitative Sobolev Approximation Bounds for Neural Operators with Empirical Validation on Burgers Equation ​
Author: Nicole Hao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, math.FA
arXiv:2605.08170v2 Announce Type: replace Abstract: Neural operators have emerged as a powerful tool for learning mappings between infinite-dimensional function spaces. However, their approximation properties in Sobolev norms remain poorly quantified, even though these norms control both function va...
416. OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents ​
Author: Xinyu Li, Ronghui Mu, Lin Li, Tianjin Huang, Gaojie Jin
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.08876v3 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents that execute tool-augmented, multi-step tasks, where latency is a critical factor for real-world applications. Yet an overlooked threat is Reasoning-Level Denial-of-Service...
417. Key-Value Means: Transformers with Expandable Block-Recurrent Compressed Memory ​
Author: Daniel Goldstein, Navneel Singhal, Eugene Cheah
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2605.09877v5 Announce Type: replace Abstract: Recall presents a difficult choice: transformers have a linearly growing memory that slows each successive token, while linear RNNs typically have fixed costs but limited recall. We present Key-Value Means ("KVM"), a novel block-recurrence for atte...
418. Formally Verifying Analog Neural Networks Under Process Variations Using Polynomial Zonotopes ​
Author: Yasmine Abu-Haeyeh, Tobias Ladner, Matthias Althoff, Lars Hedrich
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.10474v2 Announce Type: replace Abstract: Analog neural networks are gaining attention due to their efficiency in terms of power consumption and processing speed. However, since analog neural networks are implemented as physical circuits, they are highly sensitive to manufacturing process ...
419. The Transformer as a Polar State Estimator ​
Author: Peter Racioppo
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.11007v3 Announce Type: replace Abstract: We show that the core components of the Transformer---attention, residual connections, and normalization---arise naturally from a single geometric state estimation problem. Modeling the latent state in polar coordinates naturally separates radial a...
420. GeoFlowVLM: Geometry-Aware Joint Uncertainty for Frozen Vision-Language Embedding ​
Author: Mayank Nautiyal, Li Ju, Andreas Hellander, Ekta Vats, Prashant Singh
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.13352v2 Announce Type: replace Abstract: Standard dual-encoder vision-language models that map images and text to deterministic points on a shared unit hypersphere through $\ell_2$ normalization typically expose neither \emph{aleatoric} uncertainty (cross-modal ambiguity) nor \emph{episte...
421. Margin-Adaptive Confidence Ranking for Reliable LLM Judgement ​
Author: Gaojie Jin, Yong Tao, Lijia Yu, Tianjin Huang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.15416v3 Announce Type: replace Abstract: Jung et al. (2025) introduce a hypothesis testing framework for guaranteeing agreement between large language models (LLMs) and human judgments, relying on the assumption that the model's estimated confidence is monotonic with respect to human-disa...
422. When Bits Break Recourse: Counterfactual-Faithful Quantization ​
Author: Chaymae Yahyati, Ismail Lamaakal, Khalid El Makkaoui, Ibrahim Ouahbi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2605.17160v3 Announce Type: replace Abstract: Model quantization is widely used to reduce memory, latency, and deployment cost, and is typically judged by whether predictive accuracy is preserved. In decision systems that provide algorithmic recourse, however, accuracy preservation is not suff...
423. Randomized Advantage Transformation (RAT): Computing Natural Policy Gradients via Direct Backpropagation ​
Author: Mingfei Sun
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.18591v2 Announce Type: replace Abstract: Natural policy gradients improve optimization by accounting for the geometry of distribution space, but their practical use is limited by the cost of estimating and inverting the Fisher matrix. We present Randomized Advantage Transformation (RAT), ...
424. An Evidence Hierarchy for Bayesian Object Classification via OSINT-Aided Heterogeneous Sensor Fusion ​
Author: Jan Nausner, Michael Hubner
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.RO
arXiv:2605.22259v2 Announce Type: replace Abstract: Heterogeneous sensor fusion is vital for detecting, localizing, and classifying CBRNE threats. However, individual sensors are often only capable of detecting a subset of relevant threats with varying reliability or can even provide only indirect t...
425. CogAdapt: Adapting Clinical ECG Foundation Models for Wearable Cognitive Load Assessment ​
Author: Amir Mousavi, Erfan Nourbakhsh, Mohammad Sadegh Sirjani, Mimi Xie, Rocky Slavin, Leslie Neely, John Davis, John Quarles
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC
arXiv:2605.22774v5 Announce Type: replace Abstract: Assessing cognitive load continuously and at low latency would help adaptive human-computer interaction, but it remains hard because labeled data are scarce and models generalize poorly across subjects. Recent ECG foundation models, pre-trained on ...
426. MambaGaze: Bidirectional Mamba with Explicit Missing Data Modeling for Cognitive Load Assessment from Eye-Gaze Tracking Data ​
Author: Amir Mousavi, Mohammad Sadegh Sirjani, Erfan Nourbakhsh, Mimi Xie, Rocky Slavin, Leslie Neely, John Davis, John Quarles
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC
arXiv:2605.22775v3 Announce Type: replace Abstract: Real-time cognitive load assessment from eye-tracking signals could enable adaptive human-centered AI in safety-critical applications such as driver vigilance monitoring or automated flight deck assistance, yet two challenges persist: handling freq...
427. Sign-Aware Gated Sparse Autoencoders: Modeling Anticorrelated Features with Bi-Jump-ReLU Activations ​
Author: Bartosz Wieciech, Zmnako Awrahman, Marcin Czelej, Victor Hugo Jaramillo Velasquez, Wioletta Stobieniecka
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.28149v3 Announce Type: replace Abstract: Sparse Autoencoders (SAEs) extract interpretable features from Large Language Model activations, but standard variants enforce non-negative latents, so a bidirectional semantic axis (e.g., "pressure too high" vs. "pressure too low") must be split a...
428. DisjunctiveNet: Neural Symbolic Learning via Differentiable Convexified Optimization Layers ​
Author: Shraman Pal, Can Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2605.30456v3 Announce Type: replace Abstract: Many learning tasks in science and engineering are characterized by sparse datasets, which limits the effectiveness of purely data-driven approaches. At the same time, these problems are often accompanied by rich domain knowledge derived from physi...
429. Spectral Reach: Understanding Neural Scaling as Progress into the Spectral Tail ​
Author: Konstantin Nikolaou, Jonas Scheunemann, Sven Krippendorf, Samuel Tovey, Christian Holm
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2605.31244v2 Announce Type: replace Abstract: Neural scaling laws describe predictable power-law relationships between model size, dataset size, compute, and performance. While these laws guide the development of modern foundation models, the mechanisms underpinning them remain poorly understo...
430. Plug-and-Play Guidance for Discrete Diffusion Models via Gradient-Informed Logit Correction ​
Author: Hongkun Dou, Zike Chen, Fengji Li, Hongjue Li, Yue Deng
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.06303v2 Announce Type: replace Abstract: Controllable generation with discrete diffusion models is often hindered by high computational overhead or the need for retraining. In this paper, we present \underline{\textbf{G}}radient-\underline{\textbf{I}}nformed \underline{\textbf{L}}ogit \un...
431. Using Seismic Statistical Features and VQ-VAE to Improve Spatiotemporal Seismicity Predictability ​
Author: Wei Quan, Denise Gorse
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, physics.geo-ph
arXiv:2606.10069v3 Announce Type: replace Abstract: In this paper we build upon a previous study in which we demonstrated, using XGBoost and earthquake catalogue data from Japan and Chile, that a set of 60 seismic statistical features (SSFs) had much greater predictive value than a set of 428 generi...
432. CARE: Context-Aware Ranking Evolution with Executable Scoring Programs for Budgeted Reaction Optimization ​
Author: Guanyu Liu, Weiyi Kong, Chao Tang, Zeyu Wang, Boer Zhang, Baiqing Li, Peiyu Zhang, Tianyu Shi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.14581v4 Announce Type: replace Abstract: High-throughput experimentation can evaluate many reaction conditions, yet combinatorial condition spaces still exceed the available experiment budget. This makes experiment selection a sequential decision problem: each new condition must be chosen...
433. EnvShip: A Unified Framework for Context-Aware and Cross-Region Vessel Trajectory Forecasting ​
Author: Kun Ma, Qilong Han, Chengjing Song, Jingzheng Yao, Hao Wang, Changmao Wu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.15240v2 Announce Type: replace Abstract: Accurate vessel trajectory forecasting is essential for maritime situational awareness, navigation safety, traffic management, and autonomous navigation. Public Automatic Identification System (AIS) archives have enabled extensive research in this ...
434. Distilling Drifting Transformers with Representation Autoencoders ​
Author: Jiawei Zhang, Mengfei Xia, Gen Li, Yuantao Gu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.15553v2 Announce Type: replace Abstract: Despite the significant training acceleration and promising performance, Representation Autoencoders (RAEs) are mainly criticized for poor distillation effectiveness. In this work, we argue that RAE is competent at high-quality one-step generation....
435. Entropy-Gated Latent Recursion ​
Author: Soham Bhattacharjee, Dushyant Singh Chauhan, Salem Lahlou, Martin Takac, Nils Lukas
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.16620v4 Announce Type: replace Abstract: Inference-time scaling has become the dominant lever for improving language-model reasoning, but existing methods derive rollout diversity from a single source: stochastic token-level sampling. We argue that this single-axis sampling space is funda...
436. Towards Anomaly Detection on Relational Data ​
Author: Shiyuan Li, Yunfeng Zhao, Yue Tan, Qingfeng Chen, Yixin Liu, Shirui Pan
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.18621v2 Announce Type: replace Abstract: Relational databases are widely used for managing structured data in real-world systems. Detecting anomalies from such relational data is crucial for identifying fraud, risks, and abnormal behaviors, yet remains under-explored. The key challenges l...
437. DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams ​
Author: Cong Wan, Zeyu Guo, Zijian Cai, Jiangyang Li, SongLin Dong, Lin Peng, Xiangyang Luo, Zhiheng Ma, Yihong Gong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.21337v2 Announce Type: replace Abstract: Raw multimodal streams are abundant but noisy, redundant, and unaligned with any particular training objective. Turning them into supervision today means either brittle heuristics or repeatedly querying a proprietary vision-language model, a cost t...
438. HERALD: High-Throughput Block Diffusion LLM Serving via CPU-GPU Cooperative KV Cache Retrieval ​
Author: Omin Kwon, Doyeon Kim, Jongseok Park, Seung Yul Lee, Ion Stoica, Jae W. Lee
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.21633v2 Announce Type: replace Abstract: The KV cache dominates GPU memory in long-context LLM serving, crowding out batch capacity and leaving GPU compute idle. Offloading the cache to CPU DRAM restores capacity, but the limited PCIe bandwidth forces state-of-the-art offloading systems t...
439. Prefix-Guided On-Policy Distillation: Mining Golden Trajectories from Rollouts ​
Author: Qingfei Zhao, Huan Song, Shuyu Tian, Jiawei Shao, Xuelong Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.21994v2 Announce Type: replace Abstract: On-policy distillation (OPD) improves reasoning models by applying dense teacher supervision on student-sampled trajectories. However, scaling OPD to long-horizon reasoning exposes a reliability and efficiency problem: standard OPD assigns every ca...
440. DRIFT: Difficulty Routing Self-DIstillation with Rhythm-Gated Exploration and Success BuFfer Training ​
Author: Yiwei Liu, Haoning Wang, Haisen Luo, Dan Liu, Junxi Yin, Haotian Wang, Lei Zhang, Xiaoyu Tian, Shuaiting Chen, Yuansheng Song, Baoyan Guo, Xiongfei Yan, Bolan Yang, Chengwei Liu, Ming Cui, Jiong Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.30345v3 Announce Type: replace Abstract: Enabling large language models to achieve stable self-improvement without external expert supervision remains a central challenge in complex reasoning tasks. Existing self-distillation and reinforcement learning methods lack explicit mechanisms for...
441. Contextual Slate GLM Bandits with Limited Adaptivity ​
Author: Tanmay Goyal, Sukruta Prakash Midigeshi, Gaurav Sinha
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2606.31449v2 Announce Type: replace Abstract: We investigate the contextual slate bandit problem with generalized linear rewards under limited adaptivity. At each round, the learner is presented with $N$ sets of items, where each item is represented by a $d$-dimensional feature vector. The lea...
442. ECHO: Prune To Act, Trace To Learn With Selective Turn Memory In Agentic RL ​
Author: Zijun Xie, Binbin Zheng, Enlei Gong, Jihua Liu, Yuyang You, Lingfeng Liu, Jiayao Tang, Guanqun Zhao, Aoqi Hu, Zeyu Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.31650v5 Announce Type: replace Abstract: Long-horizon language agents must repeatedly interact with tools, accumulate evidence, and make decisions under bounded context windows. Context-management methods make such rollouts feasible by simplifying past interactions through deletion, foldi...
443. x-Prediction Is All You Need:Training-Free Accelerated Generation via Endpoint Decodability ​
Author: Xin Peng, Ang Gao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.06114v4 Announce Type: replace Abstract: Diffusion and flow matching models generate high-quality samples, but their ODE samplers often need tens to hundreds of neural function evaluations (NFEs). This remains a practical challenge for released checkpoints, since many accelerators require...
444. Dimensionality Reduction Meets Network Science: Sensemaking on UMAP's kNN Graph ​
Author: Duen Horng Chau, Donghao Ren, Fred Hohman, Dominik Moritz
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DS, cs.HC
arXiv:2607.08746v3 Announce Type: replace Abstract: While UMAP is widely used for exploring high-dimensional data, typical workflows focus on its lower-dimensional embedding, largely overlooking the rich k-nearest-neighbor (kNN) graph that UMAP constructs internally. This graph encodes the data mani...
445. From Direction to Magnitude: How Multimodal Instruction-Tuning Reorganizes the Geometric Encoding of Identity-Specifying Prompts in Transformer Hidden States ​
Author: Jorge A. Castillo, Marco Torres Y'evenes, Juan Carlos Lanas
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.09842v2 Announce Type: replace Abstract: We investigate whether identity-specifying system prompts produce statistically distinguishable geometric fingerprints in the hidden-state trajectories of four open-weight transformer language models spanning four post-training regimes: no training...
446. Beyond Scaffold Splits: Structural-Frontier Evaluation Reveals Hidden Failures in ADMET Models ​
Author: Jiacheng Zheng, Chang Guo, Zixuan Wang, Xinyu Liu, Hao Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2607.10729v2 Announce Type: replace Abstract: Molecular property models are commonly evaluated by holding out Bemis-Murcko scaffolds, yet a scaffold identifier is only one notion of chemical unfamiliarity. We introduce a label-free structural-frontier split that reserves the sparsest and most ...
447. DAG-FM: A Foundation Model for Causal Discovery under Heterogeneous Causal Mechanisms ​
Author: Yikang Chen, Zhengkang Guan, Haoyuan Qian, Xingxuan Zhang, Peng Cui, Yi Yang, Fei Wu, Kun Kuang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.11510v2 Announce Type: replace Abstract: Causal discovery from observational tabular data remains fundamentally challenging, primarily due to the heterogeneity of underlying causal mechanisms and the high-dimensional combinatorial search space of Directed Acyclic Graphs (DAGs). In this pa...
448. From Preimage Search To Source-Grounded Feature Inversion ​
Author: Kaixiang Shu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.12526v2 Announce Type: replace Abstract: Interpreting a neural network requires understanding what its internal features extract from a particular input. Feature inversion seeks to express a selected feature in the input domain, but canonical iterative methods search for an input whose re...
449. Reassessing Muon for Matrix Factorization ​
Author: Ali Parviz, Gal Mishne, Alex Cloninger
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13246v2 Announce Type: replace Abstract: Muon has recently emerged as a strong optimizer for large-scale deep learning, where it reshapes gradient updates through approximate orthogonalization and has been reported to outperform Adam and AdamW in large language model training. Its empiric...
450. When Does Muon Help Agentic Reinforcement Learning? ​
Author: Kai Ruan, Jinghao Lin, Zihe Huang, Ziqi Zhou, Qianshan Wei, Xuan Wang, Hao Sun
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16169v4 Announce Type: replace Abstract: Muon is competitive with AdamW in large-scale pre-training, but its operating regime in reinforcement-learning post-training remains unclear. We map this regime on ALFWorld, a sparse-reward agentic benchmark, using three group-based objectives and ...
451. Dimension-Calibrated Unexplained Mass: An Interpretable GMM Drift Statistic that Matches Kernel Two-Sample Tests ​
Author: Behnam Asadi
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16811v3 Announce Type: replace Abstract: Drift detectors that work tend not to explain themselves, and drift detectors that explain themselves tend to fail in high dimension. We close that gap for Gaussian mixture models (GMMs). Fitting a GMM to normal data makes each component a named "r...
452. DynImmune-BERT: Dynamic Immune Repertoire Modeling with Neural ODE Driven Continuous Transformers ​
Author: Rong Fu, Yongtai Liu, Xiaowen Ma, Haoyu Zhao, Shuo Yin, Yiqing Lyu, Long Zhang, Wangyu Wu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17244v2 Announce Type: replace Abstract: Longitudinal T cell receptor repertoires contain signals of clonal expansion, contraction, disappearance, and reappearance after immune perturbation. Static repertoire language models usually summarize a sample as a bag of sequences, so the samplin...
453. Retrieval-Augmented Interpretable Learning: Towards Task-Specific Zero-Shot Models in Healthcare ​
Author: Sazan Mahbub, Caleb Ellington, Zhiyuan Li, Yixin Yang, Souvik Kundu, Ben Lengerich, Eric P. Xing
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17508v2 Announce Type: replace Abstract: We introduce Retrieval-Augmented Interpretable Learning (RAIL), a probabilistic meta-learning framework for zero-shot generation of task-specific interpretable models that synthesizes coefficient-space structure from natural-language task descripti...
454. CriPO: Enhancing Rubric-based RL via Self-Distillation ​
Author: Mingxuan Xia, Yuhang Yang, Chao Ye, Shuai Zhu, Shenzhi Yang, Guangcheng Zhu, Yuhang Zhang, Cheng Peng, Haobo Wang, Siqing Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18082v3 Announce Type: replace Abstract: Rubric-based RL has recently shown promise in improving LLMs on open-ended tasks. A widely recognized limitation of rubric-based RL is limited exploration: criteria that no rollout manages to satisfy (Unexplored Criteria, UC) receive no optimizatio...
455. On the Limits of Support-Preserving Alignment and Bounded Filtering ​
Author: Aryan Dutt, Rui Mao, Anupam Chattopadhyay
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18295v2 Announce Type: replace Abstract: We study whether alignment schemes that reshape a base model's output distribution, combined with bounded safety filters, can drive the probability of harmful behavior to zero in modern large language models. Recent research suggests that harmful b...
456. Physical Self-Supervised Learning: IMU Sensing without Manual Labels ​
Author: Yuyang Leng (Richard), Renyuan Liu (Richard), Shaohan Hu (Richard), Peijun Zhao (Richard), Chun-Fu Chen (Richard), Songqing Chen, Shuochao Yao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18361v2 Announce Type: replace Abstract: Deep neural networks have become a promising approach for IMU-based sensing, but their scalability is fundamentally limited by costly labeled data and poor robustness to heterogeneous devices, placements, and users. Existing unsupervised and self-s...
457. Koopman Dreamer: Spectrally Constrained Latent Dynamics for Stable World-Model Imagination ​
Author: Jiaqi Li, Xinglong Zhang, Haibin Xie, Yixing Lan, Wei Pan, Xin Xu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2607.19719v2 Announce Type: replace Abstract: Latent world models improve sample efficiency in continuous control by optimizing policies over imagined latent trajectories, but common neural transitions offer limited direct control over modal persistence and error accumulation in long rollouts....
458. CEL: Comprehensive Counterfactual Explanations Library and Benchmark ​
Author: Oleksii Furman, {\L}ukasz Lenkiewicz, Marcel Musia{\l}ek, Maciej Zi\k{e}ba
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.22045v2 Announce Type: replace Abstract: Counterfactual explanations are a prominent approach in explainable artificial intelligence (xAI), providing actionable guidance on what input changes would alter a model's prediction to a desired outcome. While early methods primarily focused on m...
459. Understanding Machine Unlearning Through the Lens of Mode Connectivity ​
Author: Jiali Cheng, Hadi Amiri
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.23970v2 Announce Type: replace Abstract: Machine Unlearning aims to remove undesired information from trained models without full retraining from scratch. Despite recent progress, the loss landscape and optimization geometry of unlearning are poorly understood. In this paper, we study mac...
460. What EEG Foundation Models Encode: Dataset Identity and a Negative-Control Suite for Clinical Benchmarks ​
Author: Marzieh Zare
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2607.24519v2 Announce Type: replace Abstract: Pretrained EEG foundation models are proposed for clinical decoding, but whether reported gains transfer across populations or survive negative controls is unclear. We benchmark LaBraM, EEGMamba, CBraMod, REVE, LEAD, BENDR, and BIOT on five clinica...
461. Learned, Relied Upon, or Necessary? Separating Checkpoint Dependence from Task-Level Value in Sheaf GNNs ​
Author: Yi Liu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.25387v2 Announce Type: replace Abstract: Learned restriction maps in sheaf graph neural networks are often treated as proof that the model has discovered useful edge geometry. That conclusion does not follow from parameter movement or from a post-hoc ablation: both can show how one checkp...
462. Spend Experts Where You Are Unsure: Confidence-Adaptive Routing for Mixture-of-Experts LoRA ​
Author: Tom Saliencro, Rohan Desai, Priya Nair, Maya Lindqvist, Daniel Whitmore
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26052v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) variants of Low-Rank Adaptation (LoRA) route every token to a fixed number of experts $k$. Tokens differ in how uncertain the model is about them, so a single k over-spends on easy tokens and under-serves hard ones. We obse...
463. Weak-to-Strong On-Policy Distillation ​
Author: Fangxu Yu, Weijia Xu, Michael Xu, Tianyi Zhou, Zinan Lin
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26246v2 Announce Type: replace Abstract: On-policy distillation (OPD), which aligns a student with the teacher's token-level distribution on the student's own rollouts, is an effective paradigm for transferring capabilities across LLMs. Prevailing approaches assume a teacher at least as c...
464. The Convergence Behavior of Adam under Heavy-Tailed Noise ​
Author: Yijiang Pang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27383v2 Announce Type: replace Abstract: We establish the first convergence guarantees for the plain vector-form Adam optimizer under heavy-tailed stochastic noise. While several Adam variants are known to achieve optimal iteration complexity in bounded-variance nonsmooth nonconvex optimi...
465. Compliance2LoRA: Personalizable On-Demand Safety Alignment on Arbitrary Policy Subsets via Hypernetwork-Generated LoRA Adapters ​
Author: Pankayaraj Pathmanathan, Furong Huang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27594v2 Announce Type: replace Abstract: Post-training alignment in large reasoning models (LRMs) has significantly improved their adaptability to diverse safety compliance settings. However, as LRMs personalization for downstream users takes center stage, the demand for varying levels of...
466. Real-Time Hard Peak Age-of-Information Safety with No-Regret Learning ​
Author: Wentao Zhang, Wentao Mo
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27626v2 Announce Type: replace Abstract: Safety-critical IoT systems such as industrial closed-loop control, V2X coordination, and remote teleoperation require every sensor's peak Age of Information (peak AoI, also abbreviated PAoI) to stay below a hard per-slot deadline, not merely an av...
467. Hierarchical Multilevel Monte Carlo for Order-Optimal Neural Actor-Critic in Average-Reward CMDPs ​
Author: Ankur Naskar, Vaneet Aggarwal
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28390v2 Announce Type: replace Abstract: Constrained Markov Decision Processes (CMDPs) provide a natural framework for reinforcement learning in safety-critical applications, where agents maximize long-term reward while satisfying long-term constraints. Although primal-dual actor-critic m...
468. Kohn-Sham Spectral Embedding on Sparse Graphs at the Nishimori Temperature for Image Classification ​
Author: V. S. Usatyuk, D. A. Sapozhnikov, S. I. Egorov
Published: 8/4/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.IT, math.IT
arXiv:2607.28428v2 Announce Type: replace Abstract: We propose Kohn-Sham Spectral Embedding (KSSE), an energy-based model replacing the dense classifier of convolutional neural networks with a sparse-graph spectral embedding evaluated at the Nishimori temperature of an associated Random-Bond Ising M...
469. Simplified Quadratic Gradient: A Unified Framework Bridging Gradient Descent and Newton-Type Methods by Synthesizing Hessians and Gradients ​
Author: John Chiang
Published: 8/4/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2209.03282v5 Announce Type: replace-cross Abstract: Accelerating the convergence of second-order optimization, particularly Newton-type methods, remains a pivotal challenge in algorithmic research. In this paper, we extend previous work on the \textbf{Quadratic Gradient (QG)} and rigorously va...
470. Neural Born Series Operator for Biomedical Ultrasound Computed Tomography ​
Author: Zhijun Zeng, Yihang Zheng, Youjia Zheng, Yubing Li, Zuoqiang Shi, He Sun
Published: 8/4/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2312.15575v2 Announce Type: replace-cross Abstract: Ultrasound Computed Tomography (USCT) provides a radiation-free option for high-resolution clinical imaging. Despite its potential, the computationally intensive Full Waveform Inversion (FWI) required for tissue property reconstruction limits...
471. OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset ​
Author: Allen Roush, Yusuf Shabazz, Arvind Balaji, Peter Zhang, Stefano Mezza, Markus Zhang, Sanjay Basu, Sriram Vishwanath, Mehdi Fatemi, Ravid Shwartz-Ziv
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2406.14657v4 Announce Type: replace-cross Abstract: We introduce OpenDebateEvidence, a comprehensive dataset for argument mining and summarization sourced from the American Competitive Debate community. This dataset includes over 3.5 million documents with rich metadata, making it one of the m...
472. Information-Theoretic Foundations for Machine Learning ​
Author: Hong Jun Jeon, Benjamin Van Roy
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2407.12288v5 Announce Type: replace-cross Abstract: The progress of machine learning over the past decade is undeniable. In retrospect, it is both remarkable and unsettling that this progress was achievable with little to no rigorous theory to guide experimentation. Despite this fact, practiti...
473. NetDiff: Graph Diffusion with Improved Global Capabilities to Generate and Update Mobile Network Topologies ​
Author: F'elix Marcoccia, Victor Fagoo, Gilles Monzat, C'edric Adjih, Thomas Watteyne, Paul M"uhlethaler
Published: 8/4/2026, 4:00:00 AM
Categories: cs.SI, cs.LG, cs.NI
arXiv:2410.08238v2 Announce Type: replace-cross Abstract: We introduce NetDiff, a node-conditioned denoising diffusion model that generates directional link topologies and a two-slot transmit/receive parity for mobile ad hoc networks. Directional antennas can yield high throughput but require global...
474. Improved convergence rate of kNN graph Laplacians: differentiable self-tuned affinity ​
Author: Xiuyuan Cheng, Yixuan Tan, Nan Wu
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2410.23212v3 Announce Type: replace-cross Abstract: In graph-based data analysis, $k$-nearest neighbor ($k$NN) graphs are widely used due to their adaptivity to local data densities. Allowing weighted edges in the graph, the kernelized graph affinity provides a more general type of $k$NN graph...
475. Transfer Learning of CATE with Kernel Ridge Regression ​
Author: Seok-Jin Kim, Hongjie Liu, Molei Liu, Kaizheng Wang
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML
arXiv:2502.11331v4 Announce Type: replace-cross Abstract: The proliferation of data has sparked significant interest in leveraging findings from one study to estimate treatment effects in a different target population without direct outcome observations. However, the transfer learning process is fre...
476. A Survey of Circuit Foundation Model: Foundation AI Models for VLSI Circuit Design and EDA ​
Author: Wenji Fang, Jing Wang, Yao Lu, Shang Liu, Yuchao Wu, Yuzhe Ma, Zhiyao Xie
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2504.03711v3 Announce Type: replace-cross Abstract: Artificial intelligence (AI)-driven electronic design automation (EDA) techniques have been extensively explored for VLSI circuit design applications. Most recently, foundation AI models for circuits have emerged as a new technology trend. Un...
477. MGDFIS: Multi-scale Global-detail Feature Integration Strategy for Small Object Detection ​
Author: Yuxiang Wang, Xuecheng Bai, Chuanzhi Xu, Ying Zhou, Weidong Cai
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2506.12697v4 Announce Type: replace-cross Abstract: Small-object detection in Unmanned Aerial Vehicle (UAV) imagery requires preserving weak local evidence while using broader context to separate tiny foreground targets from cluttered backgrounds. Existing multi-scale fusion methods improve fe...
478. PB$^2$: Preference Space Exploration via Population-Based Methods in Preference-Based Reinforcement Learning ​
Author: Brahim Driss, Alex Davey, Riad Akrour
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2506.13741v2 Announce Type: replace-cross Abstract: Preference-based reinforcement learning (PbRL) has emerged as a promising approach for learning behaviors from human feedback without predefined reward functions. However, current PbRL methods face a critical challenge in effectively explorin...
479. Computational Approaches to Understanding Large Language Model Impact on Writing and Information Ecosystems ​
Author: Weixin Liang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.HC, cs.LG
arXiv:2506.17467v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown significant potential to change how we write, communicate, and create, leading to rapid adoption across society. This dissertation examines how individuals and institutions are adapting to and engaging ...
480. Self-Correction Bench: Uncovering and Addressing the Self-Correction Blind Spot in Large Language Models ​
Author: Ken Tsui
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2507.02778v3 Announce Type: replace-cross Abstract: Although large language models (LLMs) have transformed AI, they still make errors and follow unproductive reasoning paths. Self-correction is vital for safety-critical applications, but studying it requires disentangling activation failure fr...
481. Machine-Precision Prediction of Low-Dimensional Chaotic Systems from Noise-Free Data ​
Author: Christof Sch"otz, Niklas Boers
Published: 8/4/2026, 4:00:00 AM
Categories: nlin.CD, cs.LG, math.DS
arXiv:2507.09652v2 Announce Type: replace-cross Abstract: Low-dimensional chaotic systems such as the Lorenz-63 model are commonly used to benchmark system-agnostic methods for learning dynamics from data. This study shows that learning from noise-free observations in such systems can be achieved up...
482. From Global to Local: A Scalable Benchmark for Local Posterior Sampling ​
Author: Rohan Hitchcock, Jesse Hoogland
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2507.21449v2 Announce Type: replace-cross Abstract: Degeneracy is an inherent feature of the loss landscape of neural networks, but it is not well understood how stochastic gradient MCMC (SGMCMC) algorithms interact with this degeneracy. In particular, existing global convergence guarantees fo...
483. Barron Space Representations for Elliptic PDEs with Homogeneous Boundary Conditions ​
Author: Ziang Chen, Liqiang Huang
Published: 8/4/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.AP
arXiv:2508.07559v3 Announce Type: replace-cross Abstract: We study the complexity of approximating high-dimensional second-order elliptic PDEs with homogeneous boundary conditions on the unit hypercube using Barron spaces. Under suitable Barron assumptions on the coefficients and forcing term, we pr...
484. Conditional Deep Levy Models for Exotic Derivatives: History-Aware Path Generation and P-Q Payoff Diagnostics ​
Author: Helin Zhao, Junchi Shen
Published: 8/4/2026, 4:00:00 AM
Categories: q-fin.PR, cs.LG, q-fin.RM
arXiv:2509.13374v2 Announce Type: replace-cross Abstract: We develop and audit a history-aware financial path generator based on Denoising Levy Probabilistic Models (DLPMs) for conditional equity-index path generation. The model combines symmetric alpha-stable diffusion noise with a conditional U-Ne...
485. A Semiparametric Discrete Hawkes Model with a Collapsed Gaussian-Process Prior ​
Author: Trinnhallen Brisley, Gordon Ross, Daniel Paulin
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2509.21996v3 Announce Type: replace-cross Abstract: Hawkes processes are used in settings where past events increase the likelihood of future events occurring, resulting in a natural clustering structure. Traditional Hawkes process models treat events as occurring in continuous time, but in ma...
486. Inferring Relative Consequences of Mechanical Ventilation from Observational Data Using Game-Based Comparisons ​
Author: David J. Albers, Tell D. Bennett, Jana de Wiljes, George Hripcsak, Bradford J. Smith, Peter D. Sottile, J. N. Stroh
Published: 8/4/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG, math.OC
arXiv:2510.15127v3 Announce Type: replace-cross Abstract: Identifying the effects of mechanical ventilation (MV) protocols in critical care requires analyzing data from heterogeneous patient-ventilator systems in the clinical decision-making environment. Multiscale interactions among these coupled c...
487. PyDPF: A Python Package for Differentiable Particle Filtering ​
Author: John-Joseph Brady, Benjamin Cox, Yunpeng Li, V'ictor Elvira
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2510.25693v3 Announce Type: replace-cross Abstract: State-space models (SSMs) are a widely used tool in time series analysis. In the complex systems that arise from real-world data, it is common to employ particle filtering (PF), an efficient Monte Carlo method for estimating the hidden state ...
488. Efficiency vs. Alignment: Investigating Safety and Fairness Risks in Parameter-Efficient Fine-Tuning of LLMs ​
Author: Mina Taraghi, Yann Pequignot, Amin Nikanjam, Mohamed Amine Merzouk, Foutse Khomh
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2511.00382v2 Announce Type: replace-cross Abstract: Organizations increasingly adapt Large Language Models (LLMs) from public repositories such as HuggingFace to downstream tasks. Prior work shows that even fine-tuning on benign datasets can weaken safety alignment, raising a practical questio...
489. Interpretable Recognition of Cognitive Distortions in Natural Language Texts ​
Author: Anton Kolonin, Anna Arinicheva
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG
arXiv:2511.05969v2 Announce Type: replace-cross Abstract: We propose a new approach to multi-factor classification of natural language texts based on weighted structured patterns such as N-grams, taking into account the heterarchical relationships between them, applied to solve such a socially impac...
490. Benign Overfitting in Linear Classifiers with a Bias Term ​
Author: Yuta Kondo
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2511.12840v2 Announce Type: replace-cross Abstract: Overparameterized models often generalize well even when they interpolate noisy training data. This is known as benign overfitting. For linear classification, Hashimoto et al. (2025) analyzed the phenomenon under a broad class of mixture dist...
491. DCC: Data-Centric Compilation of Machine Learning Kernels for Processing-In-Memory Architectures ​
Author: Peiming Yang, Sankeerth Durvasula, Ivan Fernandez, Mohammad Sadrosadati, Onur Mutlu, Gennady Pekhimenko, Christina Giannoula
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AR, cs.DC, cs.LG, cs.PF
arXiv:2511.15503v4 Announce Type: replace-cross Abstract: High-performance Host processors can integrate Processing-In-Memory (PIM) devices, which can accelerate memory-intensive kernels of Machine Learning (ML) models, including Large Language Models (LLMs), by leveraging the large memory bandwidth...
492. Near-Optimal Reinforcement Learning for Constrained Recurrence Objectives ​
Author: Dominik Wagner, Leon Witzman, Luke Ong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2511.19849v2 Announce Type: replace-cross Abstract: Recurrence objectives, where a target region must be visited infinitely often, are a fundamental class of specifications for Markov decision processes (MDPs) and form the core of $\omega$-regular and linear temporal logic (LTL) objectives. We...
493. New York Smells: A Large Multimodal Dataset for Olfaction ​
Author: Ege Ozguroglu, Junbang Liang, Ruoshi Liu, Mia Chiquier, Michael DeTienne, Wesley Wei Qian, Alexandra Horowitz, Andrew Owens, Carl Vondrick
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2511.20544v2 Announce Type: replace-cross Abstract: While olfaction is central to how animals perceive the world, this rich chemical sensory modality remains largely inaccessible to machines. One key bottleneck is the lack of diverse, multimodal olfactory training data collected in natural set...
494. Latent Collaboration in Multi-Agent Systems ​
Author: Jiaru Zou, Ruizhong Qiu, Gaotang Li, Xiyuan Yang, Katherine Tieu, Pan Lu, Ke Shen, Hanghang Tong, Yejin Choi, Jingrui He, James Zou, Mengdi Wang, Ling Yang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2511.20639v4 Announce Type: replace-cross Abstract: Multi-agent systems (MAS) extend large language models (LLMs) from independent single-model reasoning to coordinative system-level intelligence. While existing LLM agents depend on text-based mediation for reasoning and communication, we take...
495. Beyond Noise: A Hypothesis Testing Approach to Robust Feature Selection ​
Author: Mousam Sinha, Tirtha Sarathi Ghosh, Koushik Biswas, Ridam Pal
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2511.20851v3 Announce Type: replace-cross Abstract: Feature selection remains difficult in modern high-dimensional settings, and established methods such as Boruta and Recursive Feature Elimination are either computationally costly or lack a statistically justified stopping criterion for their...
496. Orchestrating Dual-Boundaries: An Arithmetic Intensity Inspired Acceleration Framework for Diffusion Language Models ​
Author: Linye Wei, Wenjue Chen, Pingzhi Tang, Xiaotian Guo, Le Ye, Runsheng Wang, Meng Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2511.21759v2 Announce Type: replace-cross Abstract: Diffusion-based large language models (dLLMs) have recently gained significant attention for their exceptional performance and inherent potential for parallel decoding. Existing frameworks further enhance its inference efficiency by enabling ...
497. Variance-Aware Baselines and Adaptive Learning Rates for Reinforcement Learning with Verifiable Rewards ​
Author: Zixun Huang, Jiayi Sheng, Zeyu Zheng
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2511.23310v4 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective paradigm for post-training large language models, yet the design of its baselines and learning-rate schedules remains largely heuristic. This limits our underst...
498. NORi: An ML-Augmented Ocean Boundary Layer Parameterization ​
Author: Xin Kai Lee, Ali Ramadhan, Andre Souza, Gregory LeClaire Wagner, Simone Silvestri, John Marshall, Raffaele Ferrari
Published: 8/4/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.AI, cs.LG, physics.comp-ph, physics.flu-dyn
arXiv:2512.04452v3 Announce Type: replace-cross Abstract: NORi is a machine learning (ML) parameterization of ocean boundary layer turbulence that is physics-based and augmented with neural networks. NORi stands for neural ordinary differential equations (NODEs) Richardson number (Ri) closure. The p...
499. NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding ​
Author: Hyeonjeong Ha, Jinjin Ge, Bo Feng, Kaixin Ma, Gargi Chakraborty
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2601.01095v4 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved impressive progress in vision-language reasoning, yet their ability to understand temporally unfolding narratives in videos remains underexplored. True narrative understanding requires gr...
500. Gradient-based Optimisation of Modulation Effects ​
Author: Alistair Carson, Alec Wright, Stefan Bilbao
Published: 8/4/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, cs.SD
arXiv:2601.04867v3 Announce Type: replace-cross Abstract: Modulation effects such as phasers, flangers and chorus effects are heavily used in conjunction with the electric guitar. Machine learning based emulation of analog modulation units has been investigated in recent years, but most methods have...
501. Visualising Information Flow in Word Embeddings with Diffusion Tensor Imaging ​
Author: Thomas Fabian
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2601.05713v2 Announce Type: replace-cross Abstract: Understanding how large language models (LLMs) represent natural language is a central challenge in natural language processing (NLP) research. Many existing methods extract word embeddings from an LLM, visualise the embedding space via point...
502. Robust Bayesian Optimization via Tempered Posteriors ​
Author: Jiguang Li, Hengrui Luo
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ME, cs.LG
arXiv:2601.07094v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) iteratively fits a Gaussian process (GP) surrogate to accumulated evaluations and selects new queries via an acquisition function. Under local misspecification, this feedback loop can produce overconfidence precisel...
503. Physics-Informed Singular-Value Learning for Cross-Covariances Forecasting in Financial Markets ​
Author: Efstratios Manolakis, Christian Bongiorno, Rosario Nunzio Mantegna
Published: 8/4/2026, 4:00:00 AM
Categories: q-fin.ST, cs.LG, stat.ML
arXiv:2601.07687v4 Announce Type: replace-cross Abstract: Recent advances in nonlinear shrinkage yield asymptotically optimal cleaners for large covariance matrices and have been extended to empirical cross-covariances via singular-value shrinkage. However, these approaches rely on stationarity and ...
504. Searching for Quantum Effects in the Brain: A Bell-Type Test for Nonclassical Latent Representations in Autoencoders ​
Author: I. K. Kominis, C. Xie, S. Li, M. Skotiniotis, G. P. Tsironis
Published: 8/4/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, physics.bio-ph
arXiv:2601.10588v2 Announce Type: replace-cross Abstract: Whether neural information processing is entirely classical or involves quantum-mechanical elements remains an open question. Here we propose a model-agnostic, information-theoretic test of nonclassicality that bypasses microscopic assumption...
505. Turn-Based Structural Triggers: Structure-Conditioned Backdoors in Multi-Turn LLMs ​
Author: Yiyang Lu, Jinwen He, Yue Zhao, Kai Chen, Ruigang Liang, Cheng Hong, Yingjun Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2601.14340v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed as multi-turn assistants and customized through instruction tuning with project-specific training components. This practice creates a supply-chain risk when organizations reuse third-part...
506. Geometric Analysis of Token Selection in Multi-Head Attention ​
Author: Timur Mudarisov, Mikhal Burtsev, Tatiana Petrova, Radu State
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2602.01893v2 Announce Type: replace-cross Abstract: We present a geometric framework for analysing multi-head attention in large language models (LLMs). Without altering the mechanism, we view standard attention through a top-N selection lens and study its behaviour directly in value-state spa...
507. Chimera: Neuro-Symbolic Attention Primitives for Trustworthy Dataplane Intelligence ​
Author: Rong Fu, Xiaowen Ma, Kun Liu, Wangyu Wu, Ziyu Kong, Jia Yee Tan, Tailong Luo, Xianda Li, Yongtai Liu, Youjin Wang, Simon Fong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.NI, cs.AI, cs.CR, cs.LG
arXiv:2602.12851v5 Announce Type: replace-cross Abstract: Deploying expressive learning models directly on programmable dataplanes promises line-rate, low-latency traffic analysis but remains hindered by strict hardware constraints and the need for predictable, auditable behavior. Chimera introduces...
508. Nonparametric Distribution Regression Re-calibration ​
Author: 'Ad'am Jung, Domokos M. Kelen, Andr'as A. Bencz'ur
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2602.13362v2 Announce Type: replace-cross Abstract: A key challenge in probabilistic regression is ensuring that predictive distributions accurately reflect true empirical uncertainty. Minimizing overall prediction error often encourages models to prioritize informativeness over calibration, p...
509. Physics constraints and response validation in discrete-time reduced-order modeling: from idealized turbulent systems to climate dynamics ​
Author: Fabrizio Falasca, Laure Zanna
Published: 8/4/2026, 4:00:00 AM
Categories: nlin.CD, cond-mat.stat-mech, cs.LG, physics.ao-ph
arXiv:2602.13847v5 Announce Type: replace-cross Abstract: A central challenge across science and engineering is to build data-driven reduced-order models of turbulent dynamical systems that reproduce stationary statistics, predict responses to external perturbations, and remain practical for real-wo...
510. GaiaFlow: Semantic-Guided Diffusion Tuning for Carbon-Frugal Search ​
Author: Rong Fu, Jia Yee Tan, Chunlei Meng, Shuo Yin, Xiaowen Ma, Wangyu Wu, Simon Fong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2602.15423v5 Announce Type: replace-cross Abstract: As the burgeoning power requirements of sophisticated neural architectures escalate, the information retrieval community has recognized ecological sustainability as a pivotal priority that necessitates a fundamental paradigm shift in model de...
511. Missing-by-Design: Certifiable Modality Deletion for Revocable Multimodal Sentiment Analysis ​
Author: Rong Fu, Ziming Wang, Chunlei Meng, Jiekai Wu, Kangan Qian, Hao Zhang, Simon Fong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2602.16144v5 Announce Type: replace-cross Abstract: As multimodal systems increasingly process sensitive personal data, the ability to selectively revoke specific data modalities has become a critical requirement for privacy compliance and user autonomy. We present Missing-by-Design (MBD), a u...
512. A Comparative analysis of Layer-wise Representational Capacity in AR and Diffusion LLMs ​
Author: Raghavv Goel, Risheek Garrepalli, Sudhanshu Agrawal, Chris Lott, Mingu Lee, Fatih Porikli
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2603.07475v4 Announce Type: replace-cross Abstract: Autoregressive (AR) language models build representations incrementally via left-to-right prediction, while diffusion language models (dLLMs) are trained through full-sequence denoising. Although recent dLLMs match AR performance, whether dif...
513. OSMDA: OpenStreetMap-based Domain Adaptation for Remote Sensing VLMs ​
Author: Stefan Maria Ailuro (INSAIT, Sofia University "St. Kliment Ohridski"), Mario Markov (INSAIT, Sofia University "St. Kliment Ohridski"), Mohammad Mahdi (INSAIT, Sofia University "St. Kliment Ohridski"), Delyan Boychev (INSAIT, Sofia University "St. Kliment Ohridski"), Luc Van Gool (INSAIT, Sofia University "St. Kliment Ohridski"), Danda Pani Paudel (INSAIT, Sofia University "St. Kliment Ohridski")
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2603.11804v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) adapted to remote sensing rely heavily on domain-specific image-text supervision, yet high-quality annotations for satellite and aerial imagery remain scarce and expensive to produce. Prevailing pseudo-labeling p...
514. From AI Weather Prediction to Infrastructure Resilience: A Real-Time Correction-Downscaling Framework for Tropical Cyclone Impact Forecasting ​
Author: You Wu, Zhenguo Wang, Naiyu Wang
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2603.12828v2 Announce Type: replace-cross Abstract: This paper addresses a missing capability in infrastructure resilience: turning fast, global AI weather forecasts into asset-scale, actionable risk intelligence. We introduce the AI-based Correction-Downscaling Framework (ACDF), which combine...
515. MAPLE: Metadata Augmented Private Language Evolution ​
Author: Eli Chien, Yuzheng Hu, Ryan McKenna, Shanshan Wu, Zheng Xu, Peter Kairouz
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CR, cs.LG
arXiv:2603.19258v2 Announce Type: replace-cross Abstract: Differentially private (DP) fine-tuning of large language models (LLMs) requires massive compute and full model access, which rules out state-of-the-art proprietary APIs for general users. Generating DP synthetic data offers a practical worka...
516. The No-Clash Teaching Dimension is Bounded by VC Dimension ​
Author: Jiahua Liu, Benchong Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT
arXiv:2603.23561v4 Announce Type: replace-cross Abstract: In the realm of machine learning theory, to prevent unnatural coding schemes between teacher and learner, No-Clash Teaching Dimension was introduced as provably optimal complexity measure for collusion-free teaching. However, whether No-Clash...
517. Trust or Check? Understanding the (Evolutionary) Dynamics of User Trust in AI Systems ​
Author: Adeela Bashir, Zhao Song, Ndidi Bianca Ogbo, Nataliya Balabanova, Martin Smit, Chin-wing Leung, Paolo Bova, Manuel Chica Serrano, Dhanushka Dissanayake, Manh Hong Duong, Elias Fernandez Domingos, Nikita Huber-Kralj, Marcus Krellner, Andrew Powell, Stefan Sarkadi, Fernando P. Santos, Zia Ush Shamszaman, Chaimaa Tarzi, Paolo Turrini, Grace Ibukunoluwa Ufeoshi, Victor A. Vargas-Perez, Alessandro Di Stefano, Simon T. Powers, The Anh Han
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA, nlin.AO
arXiv:2603.24742v2 Announce Type: replace-cross Abstract: As the capabilities and adoption of Artificial Intelligence (AI) systems grow, trust in these AI systems is an increasingly urgent concern. Much research has focused on models of AI governance and has primarily examined incentives for safe de...
518. Hierarchical Pre-Training of Vision Encoders with Large Language Model ​
Author: Eugene Lee, Ting-Yu Chang, Jui-Huang Tsai, Jiajie Diao, Chen-Yi Lee
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG
arXiv:2604.00086v2 Announce Type: replace-cross Abstract: The field of computer vision has experienced significant advancements through scalable vision encoders and multimodal pre-training frameworks. However, existing approaches often treat vision encoders and large language models (LLMs) as indepe...
519. Descending into the Modular Bootstrap ​
Author: Nathan Benjamin, A. Liam Fitzpatrick, Wei Li, Jesse Thaler
Published: 8/4/2026, 4:00:00 AM
Categories: hep-th, cs.LG, hep-ph
arXiv:2604.01275v3 Announce Type: replace-cross Abstract: In this paper, we attempt to explore the landscape of two-dimensional conformal field theories (2d CFTs) by efficiently searching for numerical solutions to the modular bootstrap equation using machine-learning-style optimization. The torus p...
520. LangFIR: Discovering Sparse Language-Specific Features from Monolingual Data for Language Steering ​
Author: Sing Hieng Wong, Hassan Sajjad, A. B. Siddique
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2604.03532v2 Announce Type: replace-cross Abstract: Large language models (LLMs) show strong multilingual capabilities, yet reliably controlling the language of their outputs remains difficult. Representation-level steering addresses this by adding language-specific vectors to model activation...
521. The Illusion of Stochasticity in LLMs ​
Author: Xiangming Gu, Soham De, Michalis Titsias, Larisa Markeeva, Petar Veli\v{c}kovi'c, Razvan Pascanu
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2604.06543v2 Announce Type: replace-cross Abstract: In this work, we demonstrate that reliable stochastic sampling is a fundamental yet unfulfilled requirement for Large Language Models (LLMs) operating as agents. Agentic systems are frequently required to sample from distributions, often infe...
522. QARIMA: A Quantum Approach To Classical Time Series Analysis ​
Author: Nishikanta Mohanty, Bikash K. Behera, Badshah Mukherjee, Pravat Dash, Giuseppe Sergioli, Roberto Giuntini
Published: 8/4/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG
arXiv:2604.08277v3 Announce Type: replace-cross Abstract: We present QARIMA, a quantum state-similarity-based reconstruction of the classical ARIMA modelling pipeline. Rather than using a quantum circuit as a standalone forecaster, QARIMA preserves ARIMA's interpretable forecasting structure while r...
523. Attention Sinks as Internal Signals for Hallucination Detection in Large Language Models ​
Author: Jakub Binkowski, Kamil Adamczewski, Tomasz Kajdanowicz
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2604.10697v2 Announce Type: replace-cross Abstract: Large language models frequently exhibit hallucinations: fluent and confident outputs that are factually incorrect or unsupported by the input context. While recent hallucination detection methods have explored various features derived from a...
524. Tail-Aware Information-Theoretic Bounds for LLM Alignment under Heavy-Tailed Rewards ​
Author: Huiming Zhang, Binghan Li, Wan Tian, Qiang Sun
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, math.PR, math.ST, stat.TH
arXiv:2604.10727v2 Announce Type: replace-cross Abstract: Classical information-theoretic learning bounds typically rely on KL mutual information and moment-generating-function (MGF) arguments, which are well matched to bounded or sub-Gaussian losses but can be ineffective when losses or rewards are...
525. QShield: Securing Neural Networks Against Adversarial Attacks using Quantum Circuits ​
Author: Navid Azimi, Aditya Prakash, Yao Wang, Li Xiong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CV, cs.LG, quant-ph
arXiv:2604.10933v2 Announce Type: replace-cross Abstract: Deep neural networks remain highly vulnerable to adversarial perturbations, limiting their reliability in security- and safety-critical applications. To address this challenge, we introduce QShield, a modular hybrid quantum-classical neural n...
526. AdaDINO: Context-Adaptive DINO-Distilled Vision Foundation Models for Efficient Open-Vocabulary Edge Inference ​
Author: Yiwei Zhao, Yi Zheng, Huapeng Su, Jieyu Lin, Stefano Ambrogio, Cijo Jose, Michael Ramamonjisoa, Patrick Labatut, Barbara De Salvo, Chiao Liu, Phillip B. Gibbons, Ziyun Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2604.15622v3 Announce Type: replace-cross Abstract: Always-on contextual AI runs language-aligned vision foundation models (VFMs) on edge devices, where the on-device model is the dominant continuous compute cost under strict latency and power limits. Due to an observed low-frequency shift in ...
527. New non-Euclidean neural quantum states from hyperbolic Lorentz recurrent architectures ​
Author: H. L. Dao
Published: 8/4/2026, 4:00:00 AM
Categories: quant-ph, cond-mat.dis-nn, cs.LG
arXiv:2604.24337v3 Announce Type: replace-cross Abstract: In this work, we construct new non-Euclidean neural quantum states (NQS) based on hyperbolic Lorentz recurrent architectures (RNN/GRU). These constructions, together with the Poincare RNN NQS also newly constructed here, extend the class of p...
528. FitText: Evolving Agent Tool Ecologies via Memetic Retrieval ​
Author: Kyle Zheng, Han Zhang, Renliang Sun, Chenchen Ye, Wei Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.IR, cs.LG, cs.MA
arXiv:2605.02411v3 Announce Type: replace-cross Abstract: Efficient reasoning is not only a matter of shortening an answer trace; for tool-using agents, it also depends on whether the agent is reasoning over the right action space. As API ecosystems scale to tens of thousands of endpoints, the seman...
529. Local-Time Riemannian Score Matching on the Quantum Pure-State Manifold ​
Author: Jian Xu, Wei Chen, Shigui Li, Chao Li, Delu Zeng, John Paisley, Qibin Zhao
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2605.03573v4 Announce Type: replace-cross Abstract: Score-based diffusion can be defined intrinsically on the manifold of quantum pure states, $\mathbb{CP}^{d-1}$ with the Fubini--Study metric, but no closed-form transition density is available, so the score must be supervised by a local-time ...
530. Structured Recurrent Mixers for Massively Parallelized Sequence Generation ​
Author: Benjamin L. Badger
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2605.08696v4 Announce Type: replace-cross Abstract: Over the last two decades, language modeling has experienced a shift from the use of predominantly recurrent architectures that process tokens sequentially during training and inference to non-recurrent models that process sequence elements i...
531. Optimal Regret for Single Index Bandits ​
Author: Devdan Dey, Sujoy Bhore, Avishek Ghosh
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2605.09454v3 Announce Type: replace-cross Abstract: We study the $\textit{single-index bandit}$ problem, where rewards depend on an unknown one-dimensional projection of high-dimensional contexts through an unknown reward function. This model extends linear and generalized linear bandits to a ...
532. DRIFT: Drift-Resilient Invariant-Feature Transformer for DGA Detection ​
Author: Chaeyoung Lee, Chaeri Jung, Seonghoon Jeong
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.NI
arXiv:2605.10436v2 Announce Type: replace-cross Abstract: Domain Generation Algorithms (DGAs) evolve continuously to evade botnet detection, posing a persistent challenge for dependable network defense. While deep learning-based detectors achieve strong performance under static conditions, they suff...
533. Adaptive Kernel Density Estimation with Pre-training ​
Author: Ruitong Zhang, Ke Deng
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2605.13092v3 Announce Type: replace-cross Abstract: Density estimation in high-dimensional settings is an important and challenging statistical problem.Traditional methods based on kernel smoothing are inefficient in high dimensions due to the difficulties in specifying appropriate location-ad...
534. BCI-Based Assessment of Ocular Response Time Using Dynamic Time Warping Leveraging an RDWT-Driven Deep Neural Framework ​
Author: Shantanu Sarkar, Sai Shashank Gandavarapu, Jeff Feng, Saurabh Prasad, Reza Khanbabaie, Jose L. Contreras-Vidal
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.HC, cs.LG
arXiv:2605.14883v2 Announce Type: replace-cross Abstract: Mild traumatic brain injury (mTBI) is a prevalent condition that remains difficult to diagnose in its early stages. Oculomotor dysfunction is a well-established marker of mTBI, motivating the development of portable tools that capture both ey...
535. nASR: An End-to-End Trainable Neural Layer for Channel-Level EEG Artifact Subspace Reconstruction in Real-Time BCI ​
Author: Shantanu Sarkar, Jose L. Contreras-Vidal
Published: 8/4/2026, 4:00:00 AM
Categories: eess.SP, cs.HC, cs.LG
arXiv:2605.14941v2 Announce Type: replace-cross Abstract: Electroencephalogram (EEG) signals are highly susceptible to artifacts, resulting in a low signal-to-noise ratio, which makes extraction of meaningful neural information challenging. Artifact Subspace Reconstruction (ASR) is one of the most w...
536. Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language ​
Author: Vinayshekhar Bannihatti Kumar, Disha Makhija, Manoj Ghuhan Arivazhagan, Rashmi Gangadharaiah
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2605.15607v2 Announce Type: replace-cross Abstract: Large language models (LLMs) achieve high pass rates on code generation benchmarks, yet whether they can transfer this ability to languages absent from pretraining remains poorly understood. We introduce PyLang, a minimal imperative language ...
537. Semi-MedRef: Semi-Supervised Medical Referring Image Segmentation with Cross-Modal Alignment ​
Author: Yuchen Li, Ziru Wei, Zhen Zhao, Yi Liu, Luping Zhou
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2605.15720v2 Announce Type: replace-cross Abstract: Medical referring image segmentation (MRIS) predicts lesion masks from medical images and natural-language referring expressions, but acquiring paired pixel-level annotations and referring texts is costly. Semi-supervised learning (SSL) can a...
538. Physen-Noise2Noise: Physics-Guided Self-Supervised Defocus Deblurring with Bias Correction under Low-Light Conditions ​
Author: Ziyan Huang, Lang Wu, Hongji Wang, Yifei Liu, Dongliang Tang, Hongqiao Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, stat.ML
arXiv:2605.24590v2 Announce Type: replace-cross Abstract: Low-light, long-exposure defocus deblurring remains a challenging problem due to the simultaneous presence of severe blur and complex biased noise. Existing methods typically rely on simplified noise assumptions, which limits their effectiven...
539. A Hamiltonian-Inspired Local-Operator Ansatz for Slimming Large Language Models ​
Author: Ying Lu, Peng-Fei Zhou, Qi-Xuan Fang, Pan Zhang, Shi-Ju Ran, Gang Su
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, quant-ph
arXiv:2605.25344v2 Announce Type: replace-cross Abstract: Dense linear maps carry much of the parameter and computational burden of modern neural networks, yet their dense form leaves the organization of learned couplings implicit. Quantum many-body physics organizes exponentially large operators by...
540. Behavioural Analysis of Alignment Faking ​
Author: Nathaniel Mitrani Hadida, Rhea Karty, David Williams-King, Alan Cooney
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2605.27681v2 Announce Type: replace-cross Abstract: Alignment faking (AF) refers to a model strategically complying with a training objective to avoid behavioural modification while preserving its deployment preferences. Understanding when and why AF arises matters as models grow better at dis...
541. Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts ​
Author: Zixuan Huang, Da Chen, Kecheng Huang, Lihao Yin, Xing Li, Huiling Zhen, Mingxuan Yuan, Zili Shao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.NE, cs.DC, cs.LG, cs.PF, cs.SE, cs.SY, eess.SY
arXiv:2605.30359v2 Announce Type: replace-cross Abstract: Generating high-performance GPU kernels remains challenging due to the need for both correctness and hardware-aware optimization. While large language models (LLMs) show promise in code generation, they often fail to produce kernels that are ...
542. Enhancing Regime Shift Detection Using Unstructured Data: A Study on the Treasury Market ​
Author: Mingxuan Yi, Vidal Mehra, Jing Chen, John Cartlidge
Published: 8/4/2026, 4:00:00 AM
Categories: q-fin.CP, cs.AI, cs.LG, q-fin.ST
arXiv:2605.30363v2 Announce Type: replace-cross Abstract: Regime shifts in financial markets reorganise the joint dynamics of asset prices and macro variables, breaking any single-regime calibration. They are nonetheless hard to identify: the data signal is noisy and heavily multicollinear, while th...
543. TLA-Prover: Verifiable TLA+ Specification Synthesis via Preference-Optimized Low-Rank Adaptation ​
Author: Eric Spencer, Arslan Bisharat, Brian Ortiz, Khushboo Bhadauria, Mujtaba Nazari, TaiNing Wang, George K. Thiruvathukal, Konstantin Laufer, Mohammed Abuhamad
Published: 8/4/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG, cs.LO
arXiv:2606.06133v5 Announce Type: replace-cross Abstract: TLA+ is a formal specification language for verifying distributed systems and safety-critical protocols. Large language models (LLMs) frequently produce TLA+ specifications that fail the TLC model checker for semantic reasons. Across 25 LLMs,...
544. Priors Persist Through Suppression: A Stroop Paradigm for Lexical Override ​
Author: Han-yu Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2606.07555v4 Announce Type: replace-cross Abstract: Glossaries, technical specifications, and system prompts routinely ask language models to use familiar words in unfamiliar ways. The instruction competes with what the word already means, and even when it wins, the pretrained prior keeps oper...
545. Function-Vector Heads Are Two Populations: Writers and Cancellers in In-Context Learning ​
Author: Han-yu Wang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2606.07560v3 Announce Type: replace-cross Abstract: Function-vector (FV) heads (Todd et al., ICLR 2024) are identified by the magnitude of their causal contribution to in-context rule tasks, and the resulting top set is treated as a single functional class. We show that it holds two. Under a s...
546. HorusEye: Language as Dynamic Attention for Emergency Visual Analysis ​
Author: Armel Yara
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2606.14741v2 Announce Type: replace-cross Abstract: We introduce HorusEye, Language as Dynamic Attention for Emergency Visual Analysis. Our investigation followed five stages. The first one is benchmarking RefCOCO-Degraded, a dataset of 15,244 images (3,811 base images x 4 conditions: Clean, F...
547. Capability Provenance in Language Models: A Case Study in Social Reasoning ​
Author: Glenn Matlin, Chandreyi Chakraborty, Saehee Eom, Mika Okamoto, Rayan Castilla, Louis Jaburi, Alvin Deng, Taywon Min, Lucia Quirke, Stella Biderman, Mark Riedl
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2606.19625v2 Announce Type: replace-cross Abstract: We use training-data attribution as an interpretable tool for capability discovery, mapping which regions of the pretraining corpus support social-reasoning versus STEM-reasoning in OLMo3-7B. Training-data attribution measures how strongly ea...
548. CLQT: A Closed-Loop, Cost-Aware, Strategy-Consistent Benchmark for Diagnostic Evaluation of LLM Portfolio-Management Agents ​
Author: Bo Qu, Mingguang Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-fin.CP, q-fin.PM
arXiv:2606.29771v2 Announce Type: replace-cross Abstract: LLM agents are increasingly cast as autonomous portfolio managers, and benchmarks have moved from financial question-answering to sequential trading. Yet most still rank agents by returns over a fixed window, a weak proxy: the market path dom...
549. Freeform Preference Learning for Robotic Manipulation ​
Author: Marcel Torne, Anubha Mahajan, Abhijnya Bhat, Chelsea Finn
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2606.32027v3 Announce Type: replace-cross Abstract: Reward design remains a central bottleneck for autonomous robot policy improvement, especially in long-horizon manipulation tasks where sparse success labels provide too little signal and binary preferences collapse many competing notions of ...
550. CodeJeNN: A simple C++ neural network generator for physics applications ​
Author: Jay Arcities, Pavel Popov, Eric J Ching, Kamal Viswanath, Ryan F Johnson
Published: 8/4/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG
arXiv:2607.02746v2 Announce Type: replace-cross Abstract: Machine learning has shown speedups for numerical methods in physics applications, but integrating Python-based libraries into high-performance C++ solvers creates performance bottlenecks. We present CodeJeNN, which bridges this gap by auto-g...
551. ActionCache: Training-Free Acceleration for Vision-Language-Action Models with Action Caching and Refinement ​
Author: Ryuji Oi, Hikari Otsuka, Kosuke Matsushima, Yuki Ichikawa, Masato Motomura, Tatsuya Kaneko, Daichi Fujiki
Published: 8/4/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG
arXiv:2607.06370v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising approach for generalizable robotic manipulations. In particular, flow-matching-based VLA models have shown remarkable success due to their capability to generate precise and smoo...
552. When Top-K Misses the Decision: Tool-Call Drift in Multi-Teacher On-Policy Distillation ​
Author: Jiabin Shen, Guang Chen, Chengjun Mao
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.07050v4 Announce Type: replace-cross Abstract: Top-K teacher logits make on-policy distillation tractable, but probability mass is not decision support. In a two-teacher tool-use setting, vanilla generalized knowledge distillation raises tool-call recall but also over-calls on direct-answ...
553. MuScriptor: An Open Model for Multi-Instrument Music Transcription ​
Author: Simon Rouard, Michael Krause, Axel Roebel, Carl-Johann Simon-Gabriel, Alexandre D'efossez
Published: 8/4/2026, 4:00:00 AM
Categories: cs.SD, cs.LG
arXiv:2607.08168v2 Announce Type: replace-cross Abstract: Existing methods for automatic music transcription are often limited to single-instrument recordings or fail on complex, real music mixes. Although previous work utilizes synthetic training data, the resulting models generalize poorly, leadin...
554. Length Penalties Make Chain-of-Thought Less Monitorable ​
Author: Bryce Little
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.09786v3 Announce Type: replace-cross Abstract: To curb overthinking and reduce inference costs, researchers now train reasoning models with penalties on chain of thought length. We find that these penalties degrade monitorability. Shorter chains of thought mention misleading hints less of...
555. Tokenizing Numerical and Embedding Features for LLM RecSys ​
Author: Zhe Xu, Ankit Peshin, Chiyu Zhang, Feng Qi, Johnson Lui, Anil Ramakrishna, Justin Johnson, Carl Hu, Kaushik Rangadurai, Luke Simon
Published: 8/4/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.10016v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used as backbone architectures for recommender systems because of their strong sequence modeling and representation learning capabilities. However, most LLM-based recommenders operate primarily on...
556. Predicted Cortex Is Not a Domain-General Prior: A Matched-Control Audit of Brain-Encoding Features for Video Memorability ​
Author: Carson Rodrigues
Published: 8/4/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.16292v4 Announce Type: replace-cross Abstract: Brain-encoding foundation models predict fMRI responses to video, audio and text well enough to win the Algonauts 2025 challenge. We ask whether their predicted responses, obtained with no scanner, are a useful feature lens for a human-behavi...
557. WHALE: A Scalable Unified Model for Recommendation with Wukong-HSTU Architecture ​
Author: Renqin Cai, Dawei Sun, Yuanjun Yao, Zhiyong Wang, Velvin Fu, Maggie Zhuang, Yu Shi, Zhongnan Fang, Xuan Cao, Jing Qian, Rui Li
Published: 8/4/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG
arXiv:2607.17017v2 Announce Type: replace-cross Abstract: As scalability becomes increasingly important in recommendation modeling, recent architectures have advanced the modeling of two broad sources of ranking signals along separate paths: non-sequence features, including user, item, context, and ...
558. OTAP: Structure-Aware Optimal Transport for Evaluating Planning and Execution in Agent Trajectories ​
Author: Babak Barazandeh, Subhabrata Majumdar, George Michailidis
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.17082v2 Announce Type: replace-cross Abstract: Large language model agents solve tasks by generating trajectories that interleave planning, tool calls, and intermediate results. Current evaluation metrics reduce such a trajectory to a binary success flag, compare it against a reference by...
559. SpecFormer: Mitigating Embedding and Attention Collapse via Spectral-Aware Transformer for Recommendation ​
Author: Yu Cui, Yi Xu, Jiahao Wang, Hao Zhang, Yu Zhang, Xiaoyi Zeng, Can Wang, Jinxin Hu, Jiawei Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.24025v2 Announce Type: replace-cross Abstract: Transformer architectures have achieved remarkable success across diverse domains; however, directly applying their standard self-attention mechanism to recommendation often yields suboptimal performance, sometimes even trailing behind well-d...
560. Cross-Cohort Spectral-Temporal Dissociation in Frozen EEG Foundation-Model Representations ​
Author: Marzieh Zare
Published: 8/4/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.ET, cs.LG
arXiv:2607.24834v2 Announce Type: replace-cross Abstract: Objective. We tested whether frozen representations from five EEG foundation models support decoding of long-range temporal correlations, measured as the detrended-fluctuation-analysis (DFA) exponent of the alpha-band amplitude envelope. Appr...
561. Early Failure Prediction from Near-Anomaly Detection: A Proactive Approach ​
Author: L{'e}a Billet (LAAS, INSA Toulouse), Louise Trav{'e}-Massuy{`e}s (LAAS-DISCO, Comue de Toulouse), Elodie Chanthery (LAAS), Alexandre Gaffet
Published: 8/4/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.26704v2 Announce Type: replace-cross Abstract: Anomaly detection methods often have uncertain behavior with respect to samples near the distribution boundary, limiting their ability to anticipate future anomalies. This work introduces the concept of near-anomalies that, while not yet anom...
562. Reviewer Scores Are Not Comparable Across Research Areas in ML Peer Review ​
Author: Binyan Xu, Xilin Dai, Fan Yang, Kehuan Zhang
Published: 8/4/2026, 4:00:00 AM
Categories: cs.DL, cs.AI, cs.LG
arXiv:2607.27209v2 Announce Type: replace-cross Abstract: Peer review at ML conferences increasingly relies on reviewer scores as the primary decision instrument. As submissions have scaled from thousands to tens of thousands per year, no systematic audit has examined whether this instrument functio...
563. Group-Reflective Self-Distillation for Agentic Reinforcement Learning ​
Author: Binbin Zheng, Zijun Xie, Guanqun Zhao, Enlei Gong, Xing Ma, Xiaoliang Fu, Zeyu Chen
Published: 8/4/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.28076v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is effective for training large language model agents. However, terminal rewards provide only coarse trajectory-level supervision, leaving successful behaviors, recurring mistakes, and inc...
564. A Distributed Acoustic Sensing Dataset for Vessel Detection and Localization in Submarine Cable Protection ​
Author: Erick Eduardo Ramirez-Torres, Javier Macias-Guarasa, Daniel Pizarro, Javier Tejedor, Sira Elena Palazuelos-Cagigas, Pedro J. Vidal-Moreno, Mar'ia R. Fern'andez-Ruiz, Sonia Martin-Lopez, Miguel Gonzalez-Herraez, Roel Vanthillo
Published: 8/4/2026, 4:00:00 AM
Categories: physics.geo-ph, cs.LG
arXiv:2607.28306v2 Announce Type: replace-cross Abstract: Recent incidents of accidental damage and suspected sabotage to submarine telecommunication and power cables, particularly in the Baltic Sea, have underscored their vulnerability and the need for continuous monitoring solutions. Distributed a...
565. MoPET: Parameter-Efficient Mixture-of-Experts for Unified Medical Image Classification ​
Author: Sebastian Doerrich, Daniel W"urtinger, Francesco Di Salvo, Shyam Nandan Rai, Christian Ledig
Published: 8/4/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2607.29462v2 Announce Type: replace-cross Abstract: Adapting deep learning models to profound clinical heterogeneity typically relies on parameter-efficient fine-tuning (PEFT) to avoid the severe overfitting associated with full end-to-end network updates. Although PEFT successfully navigates ...