arXiv cs.LG - 2026-08-03 ​
247 items collected.
1. Topology-Aware Data Movement for Disaggregated GPU Inference ​
Author: Sanjeev Rao Ganjihal
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.PF
arXiv:2607.28633v1 Announce Type: new Abstract: Disaggregated LLM inference creates a datacenter networking problem that no existing system solves correctly. When prefill and decode run on separate GPU pools, the KV cache must be transferred between them. For a 70B model this is 2.6 GB per request, ...
2. Sensitivity Analysis of GRU, LSTM and Transformer Encoder in Classification of Automated Driving Systems ​
Author: Bidhya Shrestha, Christos Papadopoulos
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.28665v1 Announce Type: new Abstract: Automated driving systems (ADSs) are becoming ubiquitous. Future Software Defined Vehicles (SDVs) may be able to run multiple ADSs, both native and aftermarket such as Comma.ai's Openpilot. Monitoring systems to independently verify which automated dri...
3. Guarantees on Dynamical System Distinguishability for LLM Token Generation ​
Author: Mohamed Akrout, Dan Wilson
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DS
arXiv:2607.28667v1 Announce Type: new Abstract: Recent work has shown that classifying large language models (LLMs)' responses can be distinguished by modeling token embeddings as trajectories of a black-box dynamical system (DS) and comparing prediction residuals of two DSs. Despite the empirical s...
4. LARA: Lightweight Adapters in the Residual Stream for Composable Adaptation and Alignment ​
Author: Pascal Ekin, Hyosun Choi, Wei Jie
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28669v1 Announce Type: new Abstract: We present LARA (Lightweight Additive Residual Adaptation), a method for efficient adaptation that operates in the residual stream of a frozen model rather than in its weights. Where LoRA adds an update of low rank to weight matrices, LARA reads the hi...
5. Hierarchical Copula-Gumbel-Top-\texorpdfstring{$K$}{K} Routing: Two-Sided Dependence Control for Frozen Mixture-of-Experts at Fixed Per-Token Routing Laws ​
Author: Richard Yi Da Xu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28670v1 Announce Type: new Abstract: A stochastic Gumbel-Top-$K$ router defines, for every token of a mixture-of-experts (MoE) model, a \emph{routing law}: a distribution over ordered expert lists and mixture weights. We ask which \emph{joint} distributions over the routing choices of dif...
6. LAWFUL: Law-Aligned Witness for Faithful Use of Latents ​
Author: Kevin Chen, Kenneth W. Parker, Anish Arora
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.28672v1 Announce Type: new Abstract: When a neural network predicts a physical system accurately, has it learned the governing law as formal, structured knowledge, and if so, does the network's internal computation actually use that representation throughout the law's domain of validity? ...
7. MPP-GNN: Subject-Adaptive Community Detection for fMRI-Based Alzheimer's Disease Classification ​
Author: Yang Zhang, Xiao Zhou, Jonathan Warrell, Avram Holmes, Xuan Zhang, Mark Gerstein
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.NC
arXiv:2607.28681v1 Announce Type: new Abstract: Functional magnetic resonance imaging (fMRI) is a widely used technique for studying the brain. Recent methods that utilize graph neural networks (GNNs) for analysis of brain functional connectivity have shown great potential for the classification of ...
8. Technological Advances in Detecting and Managing Cognitive Impairment in Older Adults: Trends, Challenges, and Future Directions ​
Author: Mohammad Asif, Azizuddin Khan, Mohd Azam, Anurag Rajkumar Bombarde
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET
arXiv:2607.28687v1 Announce Type: new Abstract: As populations age, cognitive decline from mild cognitive impairment (MCI) to dementia is a defining health challenge of the coming decades, yet routine assessment often misses its earliest signs. This article critically synthesizes recent technologica...
9. SEDR-Seq2P: A Lightweight Dilated Residual Sequence-to-Point Network for Multi-Task Industrial NILM ​
Author: Hatem Haddad, Feres Jerbi, Issam Smaali
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.28693v1 Announce Type: new Abstract: Industrial NILM remains challenging because measurement noise and widespread concurrent machine operation reduce the generalization of models tuned on residential data. This work adopts a one-to-many, multi-task disaggregation setting, in which a singl...
10. Predicting Steel Fatigue Life from Micrographs Using Physics-Informed Deep Learning ​
Author: Aryuemaan Kumar Chowdhury
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2607.28695v1 Announce Type: new Abstract: Here is the plain text version optimized for arXiv's submission form. Custom macros (like \CV and \SI) have been converted to standard text/math so they render correctly on the webpage: Evaluating the fatigue life of structural steels conventionally re...
11. Mitigating Class-Tail Undercoverage in Medical Vision-Language Models under Clinical Shift ​
Author: Mushir Akhtar, M. Tanveer
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.28696v1 Announce Type: new Abstract: Medical vision-language models (VLMs) can retain high observed marginal coverage after clinical shift while substantially under-covering an individual disease class. The affected class varies with acquisition protocol and backbone geometry, so source p...
12. Flow Matching with Missing Data ​
Author: Fairoz Nower Khan, Nabuat Zaman Nahim, Peizhong Ju
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28698v1 Announce Type: new Abstract: Flow matching assumes fully observed training data, which many real-world applications rarely provide. We propose Missing-Data Flow Matching, which treats the missing coordinates of training samples as latent variables and averages the flow matching lo...
13. MMFGU: Multimodal Federated Graph Unlearning ​
Author: Haodong Lu, Zekai Chen, Weiwei Ji, Shihao Li, Xunkai Li, Xun Wu, Yinlin Zhu, Rong-Hua Li
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28708v1 Announce Type: new Abstract: Multimodal federated graph learning enables clients to collaboratively train graph models over structural, textual, and visual signals without sharing private local data. However, the presence of heterogeneous multimodal content also makes unlearning r...
14. Mirror Learning ​
Author: Yunpeng Liu, Matthew Niedoba, Oluwanifemi A. Adekanye, Jason Yoo, Yingchen He, Berend Zwartsenberg, Frank Wood
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.RO
arXiv:2607.28737v1 Announce Type: new Abstract: We investigate imitation learning through the lens of third-person observation and propose a framework for mirror learning: acquiring actionable policies from passive observation. While behavior cloning (BC) excels under dense, well-aligned first-perso...
15. TAGTorch: A PyTorch Library for Geometry, Topology, and Symmetry-Aware Machine Learning ​
Author: Brendan Kennedy, Tegan Emerson, Gregory Roek, Emilie Purvine, Henry Kvinge
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28755v1 Announce Type: new Abstract: Over the last decade, neural networks have been applied to an increasingly diverse range of applications, including data with rich geometric, topological, or symmetry-related structure. As a result, researchers have increasingly drawn inspiration from ...
16. Feature Interaction Modeling for Physics-Informed Neural Networks and Neural Operators ​
Author: Quan Gu, Hongxia Liu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28762v1 Announce Type: new Abstract: This work embeds feature interaction modules derived from factorization machines (FMs) into physics-informed neural networks (PINNs) and neural operator learning, to enhance model expressiveness for solution manifolds of parameterized partial different...
17. Representations from Pretrained Machine-Learning Interatomic Potentials as Coarse Coordinates for Material Generation and Evaluation ​
Author: Paul Hagemann, Katharina Ueltzen, Simon M"uller, Janine George, Philipp Benner
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci
arXiv:2607.28776v1 Announce Type: new Abstract: Generative machine learning is increasingly used for inorganic crystal structure generation. Most models and the corresponding evaluation approaches rely on simple forms of crystal structure representation. In this paper, we showcase the power of atom-...
18. Distilling Knowledge from Large Language Models into Lightweight Reinforcement Learning Agents for Autonomous Cyber Operations ​
Author: Konur Tholl, Fran\c{c}ois Rivest, Mariam El Mezouar, Adrian Taylor, Ranwa Al Mallah
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28826v1 Announce Type: new Abstract: Autonomous Cyber Operations (ACO) are increasingly important for defending enterprise networks as cyber threats continue to evolve in sophistication. ACO applications commonly employ Reinforcement Learning (RL) agents to learn defensive behaviors throu...
19. Hypergradient-based Bilevel Reinforcement Learning with Improved Sample Complexity ​
Author: Naman Saxena, Mudit Gaur, Vaneet Aggarwal
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.28849v1 Announce Type: new Abstract: Bilevel reinforcement learning (RL) is an important framework within the literature of RL that can be used to formalize various categories of problems, such as meta-learning, hierarchical task decomposition, and reinforcement learning from human feedba...
20. An analysis of machine learning approaches for enhancing decision-making in complex discrete choice tasks ​
Author: Sheng Lun Christine Cao, Destenie Nock, Alex Davis
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28854v1 Announce Type: new Abstract: Discrete choice modeling is a common tool used for preference elicitation during policy-making, but this is typically done through parametric models. Machine learning can push the boundaries of discrete choice modeling for policy-based preference elici...
21. Fast Rates for Swap-Agnostic Learning of Proper Losses ​
Author: Princewill Okoroafor
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28856v1 Announce Type: new Abstract: Swap-agnostic learning strengthens classical agnostic learning by allowing the comparator to select a different hypothesis on each level set of the learner's predictions. This benchmark captures prediction-dependent postprocessing, but appears to requi...
22. Adaptivity via a Parallel Architecture for Stochastic Gradient Methods Adaptivity via a Parallel Architecture for Stochastic Gradient Methods Adaptivity via a Parallel Architecture for Stochastic Gradient Methods ​
Author: Bin Fu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, math.OC
arXiv:2607.28902v1 Announce Type: new Abstract: We develop a parallel framework that assembles static gradient methods to achieve better adaptivity. A static gradient method, denoted by $\mathrm{GD}(x_0,T)$, takes as input an initial point $x_0\in\mathbb{R}^n$ and $T\in \mathbb{R}^+$ specifying the ...
23. Reflection or Re-Generation? Why LLM Revision Fails Where Human Revision Succeeds ​
Author: Yefan Tao, Gerald Friedland, Madhusudhanan Chandrasekaran, Luyang Kong
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28908v1 Announce Type: new Abstract: Reflection, the ability to revisit and revise prior reasoning, is central to how humans improve their answers. Large language models (LLMs) are increasingly prompted to "reflect," yet whether this resembles human revision remains unclear. We introduce ...
24. Gated Q-learning: Add Off-Policy Bias to Taste ​
Author: Brett Daley
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.28916v1 Announce Type: new Abstract: Multistep credit assignment is critical for sample-efficient reinforcement learning, yet managing off-policy bias in Q-learning remains a fundamental challenge. For 30 years, practitioners have been limited to a binary choice: eliminate the bias at the...
25. Learning Optimal Dynamic Matching via Graph Neural Networks ​
Author: Genta Okada, Shunya Noda, Junpei Komiyama, Akira Matsushita
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.GT, econ.TH
arXiv:2607.28925v1 Announce Type: new Abstract: Dynamic matching markets require decisions about whom to match and when: matching now yields value but removes participants who may create better future opportunities. We develop a value-based reinforcement-learning framework for this problem on finite...
26. Latent Lie-Poisson Neural Networks (LLPNNs): Discovering the motion of Lie-Poisson systems through observable data and latent dynamics ​
Author: Vakhtang Putkaradze
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, math.DS, math.OC
arXiv:2607.28939v1 Announce Type: new Abstract: Structure-preserving neural networks are essential for the long-term prediction of Hamiltonian systems from data. Many important Hamiltonian systems in mechanics and control admit symmetry reduction to Lie--Poisson systems, including rigid bodies, unde...
27. FairDiffuseVQVAE: Sampling-Time Fairness in Tabular Diffusion via Conditional Refinement of Vector-Quantized Latents ​
Author: Nitish Nagesh, Mahdi Bagheri, Amir M. Rahmani
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28945v1 Announce Type: new Abstract: Synthetic tabular data is increasingly used in privacy-preserving data sharing, data augmentation, and to mitigate downstream classifier bias. State-of-the-art tabular diffusion models such as TabDDPM and TabSyn achieve excellent distributional fidelit...
28. Shapley-Value-Based Feature Attribution for Data Masking ​
Author: Xinxue (Shawn), Qu, Francis Bilson Darku, Hong Guo
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28946v1 Announce Type: new Abstract: Despite its many benefits, widespread access to individuals' personal data also causes severe privacy concerns for consumers, companies, and policymakers. This study proposes a novel framework that adapts the Shapley-value-based feature attribution app...
29. Overcoming the Weakest-Link Effect in LLM-Driven Program Optimization via Heterogeneous Edit Recombination ​
Author: Jingwen Fu, Zhen Liu, Yuhan Liu, He Zhang, Nanning Zheng
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28947v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to solve complex problems by searching over program space, offering a general paradigm for scientific problems that can be naturally represented and solved as programs. Despite recent progress, identif...
30. Mining Verdict Boundaries for Neural Network Verification ​
Author: Jiawei Ren, Guanqin Zhang, Zhenya Zhang, Yulei Sui
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.LO, cs.SE
arXiv:2607.28954v1 Announce Type: new Abstract: Branch and Bound (BaB) aims to achieve complete verification of neural networks by adaptively partitioning the problem and applying off-the-shelf verifiers to subproblems. Its problem-splitting history can be represented as a tree, where each subproble...
31. Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates ​
Author: Weiyi He, Yuping Lin, Jiliang Tang, Yue Xing
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.28959v1 Announce Type: new Abstract: Adversarial training is one of the most effective defenses against adversarial attacks, yet the computational cost remains prohibitive at modern scales, especially for large language models (LLMs). While existing mitigation strategies, e.g., latent adv...
32. Beyond Feature and Structure Alignment: Learning Transferable Propagation Knowledge for Graph Foundation Models ​
Author: Yi Wang, Jitao Zhao, Di Jin, Dongxiao He
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28980v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) have recently emerged as a promising paradigm for enabling knowledge transfer across diverse domains. Unlike traditional graph learning methods that are typically designed for in-domain settings, GFMs aim to learn transfe...
33. SILVA Networks as Structured Implicit Layers and Vector Attractors via Dynamic Interaction Fields ​
Author: Jose Luis Lima de Jesus Silva
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2607.28989v1 Announce Type: new Abstract: Many learning problems require representations that reconcile direct input, nearby structure, and broader context. In implicit neural layers, these influences are usually absorbed into a single fixed-point update, making it hard to identify what enters...
34. Dynamics-aware identification of governing equations from sparse and noisy data ​
Author: Pongpisit Thanasutives, Yoshinobu Kawahara
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, math.DS
arXiv:2607.29036v1 Announce Type: new Abstract: Sparse identification of nonlinear dynamics (SINDy) and PDE functional identification (PDE-FIND) recover parsimonious ordinary and partial differential equations (ODEs and PDEs) from data. However, sparse and noisy temporal measurements can make deriva...
35. DFSC: Error-Controlled Differentiable Mittag-Leffler Propagation for Fractional Scientific Machine Learning ​
Author: Ning Hu, Haitao Duan, Shuqun Li, Chuyang Hu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29038v1 Announce Type: new Abstract: Fractional scientific machine learning requires numerical operators that can be differentiated, batched, accelerated, and composed with neural networks. When the dominant linear fractional evolution is known through a Mittag-Leffler propagator, repeate...
36. Learning Lookahead Lemmas for Neural Network Verification ​
Author: Liam Davis, Haoze Wu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.LO
arXiv:2607.29051v1 Announce Type: new Abstract: State-of-the-art neural network verifiers use the branch-and-bound procedure as their core solving mechanism. We introduce an inprocessing framework for neural network verification driven by the lookahead procedure. Under this framework, lookahead deri...
37. Who Wins Where? Conformal Model Comparison for Local Superiority ​
Author: Yi Zhou, Baishi Li, Xuan Yao, Ke-Wei Huang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29053v1 Announce Type: new Abstract: Standard model comparison is global, aggregating losses across the covariate space to declare a single winner. This can obscure heterogeneous performance, where different models are preferable in different regions. We introduce conformalized local mode...
38. Autonomous Repair for Multi-Agent Systems via Monte-Carlo Tree Search ​
Author: Hanxiao Lu, Tianyi Zhang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA
arXiv:2607.29055v1 Announce Type: new Abstract: Multi-agent systems (MAS) are increasingly deployed to solve complex tasks. In case of incorrect or unsatisfactory outputs, users have to manually locate agent mistakes by inspecting agent trajectories (i.e., {\em failure attribution}) and provide feed...
39. Benchmarking Frontier Large Language Models Against Official Crash Database Coding Using Police Crash Narratives ​
Author: Sudhir Bharati, Rajendra K C Khatri, Sudip Bharati
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29064v1 Announce Type: new Abstract: Police crash narratives contain information that may supplement structured crash databases, but manual review is labor-intensive and it remains unclear how well large language models (LLMs) reproduce official crash coding. This study benchmarked six fr...
40. Federated Foundation Models Fine-Tuning with Heterogeneous Compressed Clients ​
Author: Shengkun Zhu, Jinshan Zeng, Zhihua Allen-Zhao, Mayi Xu, Quanqing Xu, Wei Ren, Qiang Yang, Yang Liu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29071v1 Announce Type: new Abstract: Federated learning of foundation models faces a fundamental resource-asymmetry challenge: the institutions holding the most valuable domain-specific data cannot host billion-parameter models. Existing heterogeneous federated approaches attempt to bridg...
41. DASH-OPD: Discrepancy-Aware Switching with Hysteresis for On-Policy Distillation ​
Author: Yuchen Xia, Qianguo Sun, Chao Song, Junlong Wu, Yiyan Qi, Yunjian Xu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29078v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models on their own rollouts to reduce exposure bias. However, in multi-turn agent scenarios, early student errors can lead a trajectory away from the teacher's familiar domain. Existing curriculum learning m...
42. What Is Missing in Surgical Risk Stratification and Outcome Prediction: A Scoping Review of End-to-End Machine Learning Approaches ​
Author: Yizhi Dong, Yuhe Ke, Hairil Rizal Abdullah, Yucheng Xing, Kevan Kai Bing Teo, Ling Huang, Mengling Feng
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29090v1 Announce Type: new Abstract: Postoperative adverse events, including mortality and morbidity, remain a major global burden, many of which are preventable through early identification of high-risk patients and targeted perioperative care. Accurate risk stratification is therefore e...
43. PiDDM: Physics-Informed Differentiable Degradation Modeling for Lithium-Ion Battery State-of-Health Prediction ​
Author: Zeping Chen, Ruda Jian, Sachin Sigdel, Guoping Xiong, Jian-Xun Wang, Tengfei Luo
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29095v1 Announce Type: new Abstract: Accurate prediction of lithium-ion battery state of health (SOH) is essential for reliable energy storage operation. However, purely data-driven models may generalize poorly across cycling protocols and produce physically implausible behavior during lo...
44. StraightDP: Geometry-Aware Differential Privacy for Rectified-Flow Transformers ​
Author: Xujun Che, Depeng Xu, Xintao Wu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.CV
arXiv:2607.29100v1 Announce Type: new Abstract: Differentially private (DP) training of text-conditioned generative models suffers a utility cliff at strong privacy. We revisit this problem through the geometry of rectified flows: along the straight interpolation between noise and data, the Bayes-op...
45. Curriculum Matters: Data-Efficient Relational PFN Pretraining with Synthetic Data ​
Author: Mohammad Sadeq Abolhasani, Viswanath Ganapathy
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.DB
arXiv:2607.29120v1 Announce Type: new Abstract: Relational Prior-Data Fitted Networks (PFNs) such as RDB-PFN approximate Bayesian inference over multi-table relational databases by pretraining on millions of synthetic tasks. We investigate three intertwined questions about this paradigm. First, can ...
46. PluRel-to-RDB-PFN: Schema-Guided Synthetic Relational Pretraining ​
Author: Mohammad Sadeq Abolhasani, Viswanath Ganapathy
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29129v1 Announce Type: new Abstract: Relational Foundation Models (RFMs) require large-scale synthetic relational databases for pretraining, but existing approaches tightly couple data generation with the model training pipeline. We study whether PluRel, a general-purpose synthetic relati...
47. HERO: History-Enriched Rollout Training for Long-Horizon Autoregressive Neural Operators ​
Author: Jiaquan Zhang, Shuxu Chen, Haifan Meng, Yi Lu, Zhihan Lyu, Fan Mo, Wei Dong, Yang Yang, Chaoning Zhang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29135v1 Announce Type: new Abstract: Neural operators provide fast surrogates for time-dependent partial differential equations (PDEs) by applying a learned evolution operator recursively to its own predictions, but this autoregressive rollout feeds every prediction error back as input, s...
48. Implicit Machine Learning Force Fields Accelerate Molecular Dynamics Simulations ​
Author: Johannes Mae{\ss}, Leon Werner, J. Thorben Frank, Winfried Ripken, Martin Michajlow, Joshua Futterer, Klaus-Robert M"uller, Stefan Chmiela
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29158v1 Announce Type: new Abstract: We introduce implicit machine learning force fields (I-MLFFs), which replace explicit stacks of neural network layers with self-consistent fixed-point equations. In molecular simulations, this formulation enables intermediate representations to be reus...
49. MBDiff: Multi-view Behavior-aware Diffusion Model for Probabilistic Utility Data Imputation ​
Author: Rongchao Xu, Lin Jiang, Dahai Yu, Ximiao Li, Guang Wang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29177v1 Announce Type: new Abstract: Utility data (e.g., electricity, water, and gas consumption), collected by ubiquitous sensors and embedded devices, often contains substantial missing values due to various factors such as device failures and data transmission issues. The data missingn...
50. SERUM: State Extraction and Refinement for User Modeling ​
Author: Andy J. Phu, James Mooney, Karin de Langis, Khanh Chi Le, Dongyeop Kang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2607.29181v1 Announce Type: new Abstract: Agentic assistants capable of proactive, personalized interactions require structured models of user intent and workflow. However, building these models from raw, unstructured screen activity remains an open challenge. We present SERUM, a multi-pass fr...
51. SAF-OPD: Stable Advantage Fusion for On-Policy Distillation ​
Author: Yifan Ding, Xincheng Wei, Yoshua Y. Li, Ziheng Li, Yuquan Lu, Siyu Zhang, Dongsheng Ma, Rongxiang Weng, Xunliang Cai, Yun Chen
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29209v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) broadcasts a single response-level reward to every token, while on-policy distillation (OPD) scores each token against a stronger teacher for a dense advantage but caps performance at teacher qualit...
52. Frugal Bayesian Optimization: Scalable Surrogates for Data- and Resource-Limited Discovery ​
Author: Panagiotis Krokidas, Christoforos Rekatsinas, Vassilis Sioros, Grigorios M. Chatziathanasiou, Efi-Maria Papia, George Giannakopoulos
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph, physics.comp-ph
arXiv:2607.29225v1 Announce Type: new Abstract: Bayesian Optimization (BO) is widely adopted for data-efficient optimization in scientific and engineering applications, yet its computational cost is rarely evaluated alongside optimization performance. Here we present a systematic, compute-aware stud...
53. UniPolymer: A Unified Framework for Property Prediction, Structure Recommendation, and Evaluation in Polyimide Design ​
Author: Junquan Hu, Zhihui Wang, Peng Xu, Xinru Guo, Xintong Li, Kun Lu, Ben Fei
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29256v1 Announce Type: new Abstract: Designing polyimide structures with specific glass transition temperatures (Tg) is highly challenging. Existing methods primarily focus on target-conditioned generation, lacking an assessment of the consistency between the generated structure and the t...
54. Assessing the Generalization of Graph Neural Networks for Fault Location Across Increasing Distributed Energy Resource Penetration Levels ​
Author: Burak Karabulut, Olayiwola Arowolo, Carlo Manna, Chris Develder, Jochen L. Cremer
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2607.29293v1 Announce Type: new Abstract: Accurate fault location is critical for distribution network reliability. However, increasing distributed energy resource (DER) penetration complicates fault location due to intermittent generation and bidirectional power flows that reshape fault signa...
55. Sample Efficient Hierarchical Reinforcement Learning via Best Policy Identification ​
Author: Anders Jonsson, Emilie Kaufmann, Gianmarco Tedeschi, Lorenzo Steccanella
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29294v1 Announce Type: new Abstract: We present HBPI-UCRL, a model-based algorithm for hierarchical reinforcement learning (HRL) that learns high-level and low-level policies in parallel. HBPI-UCRL exploits the fact that a high-level transition corresponds to a multi-step transition at th...
56. Analysing User Reviews to Identify User Concerns Around Permissions in AI Apps ​
Author: Babar Shah, Faheem Ullah, Myles Watkinson, Muhammad Moiz Khalid, Tehmina Karamat Khan, Muhammad Junaid
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29343v1 Announce Type: new Abstract: Artificial intelligence is increasingly embedded in everyday software, making its integration into mobile apps inevitable. However, AI mobile app developers are not always versed in security and privacy best practices, leaving users to monitor their ow...
57. Versatile On-device Adaptation at the Edge by Unifying Few-shot, Zero-shot, Continual, and In-context Learning ​
Author: Douwe den Blanken, Martin Lefebvre, Charlotte Frenkel
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.AS
arXiv:2607.29353v1 Announce Type: new Abstract: With the ever-increasing pervasiveness of smart edge devices, the demand is growing for applications that can be tailored to users (e.g., custom keyword spotting) or patients (e.g., adaptive health monitoring). Yet, most edge devices rely on fixed infe...
58. Cross-Resolution Semantic Learning for Graph Domain Adaptation ​
Author: Yingxu Wang, Haoze Huang, Zhongkai Zheng, Shangsong Liang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29365v1 Announce Type: new Abstract: Graph Domain Adaptation (GDA) transfers predictive knowledge from labeled source graphs to unlabeled target graphs under distribution shift. Existing methods align representations or regularize graph structures, but do not explicitly model how class-di...
59. Exploring Block Anomaly Detection In HDFS Log Data Analysis ​
Author: WenYang Zhong, Tutut Herawan
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.SE
arXiv:2607.29383v1 Announce Type: new Abstract: In recent years, with the development of big data technology, increasingly more companies use HDFS for data processing and storage. As a result, the maintenance of distributed file systems has become an extremely important part of data management. As t...
60. Simulation Code Generation for Fluid Systems using Large Language Models: Benchmarking Models and Prompting Strategies ​
Author: Jan Marius St"urmer, Jascha Knack, Tobias Koch, Andreas Weinmann
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.SE
arXiv:2607.29389v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated a strong ability to generate syntactically correct code from natural-language specifications. In this study, we explore how LLMs can be harnessed to automatically translate a neutral graph representation o...
61. OnlineCache: Learning Dynamic Caching Policies with Error Correction for Efficient Diffusion Inference ​
Author: Zhikang Xie, Xichen Ye, Yifan Wu, Haoshen Yu, Li chenan, Peizhu Gong, Weizhong Zhang, Cheng Jin
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29398v1 Announce Type: new Abstract: Diffusion models have revolutionized generative tasks but incur high latency due to iterative denoising. While cache-based strategies accelerate inference by reusing intermediate features, they largely rely on static, sample-agnostic schedules. We argu...
62. ALIVE: Warnings Before Exclusion in Budgeted Multi-Source Learning ​
Author: Xiyang Zhang, Hongzhi Wang, Yuanhe Tian
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29400v1 Announce Type: new Abstract: A routing decision can be revised at the next transaction, but a latched source exclusion persists across later decisions. We ask what evidence should authorize these unequal-persistence actions when finite-population auditing and learning share a budg...
63. Explore Beyond the Boundary Using Entropic Information ​
Author: Bumgeun Park, Donghwan Lee
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29419v1 Announce Type: new Abstract: In reinforcement learning, exploration with sparse and delayed rewards presents a significant challenge due to the limited feedback available for guiding the learning process. Addressing this issue requires extensive exploration in the state space to d...
64. End-to-End Fairness Optimization with Fair Decision-Focused Learning ​
Author: Yu Wang (Xinying), Violet (Xinying), Chen
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2607.29441v1 Announce Type: new Abstract: Many real-world systems rely on predictive models to inform decisions, and fairness concerns arise in both the prediction and decision stages. We introduce end-to-end fairness optimization (E2EFO) as a unifying framework that integrates fairness across...
65. TFGformer: Multivariate Time Series Forecasting via Time-Frequency Graph Learning and Covariate Fusion ​
Author: Yu Sun, Yuan Chang, Xiaohou Shi, Yan Sun
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29459v1 Announce Type: new Abstract: Large-scale multivariate time series from heterogeneous IoT sensors demand accurate long-term forecasting for resource scheduling and predictive maintenance. While recent time series foundation models exhibit strong generalization, they rely on static ...
66. Parameter-Free Heavy-Tailed Bandits ​
Author: Gianmarco Genalti, Alberto Maria Metelli
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29460v1 Announce Type: new Abstract: Heavy-tailed distributions arise naturally in sequential decision-making problems such as financial investment, online advertising, and network management, where rare but extreme outcomes can dominate performance. Heavy-tailed bandits model online deci...
67. MolGVR: A Chemistry-Grounded Framework for Text-to-Molecule Generation ​
Author: Qian Tan, Xuanyu Zhu, Lei Jiang, Zhonghang Yuan, Chen Zhang, Yuqiang Li
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29479v1 Announce Type: new Abstract: Text-to-molecule generation is typically formulated as a one-shot sequence generation problem, where a model directly maps target descriptions to molecular representations. However, molecular descriptions often contain informative structural constraint...
68. DreamQAS: Learning a Decision-Useful World Model for VQE-Efficient Quantum Architecture Search ​
Author: Jiayang Niu, Yan Wang, Jie Li, Ke Deng, Azadeh Alavi, Muhammad Usman, Yongli Ren
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29491v1 Announce Type: new Abstract: Reinforcement-learning-based quantum architecture search (RL-QAS) repeatedly optimizes a variational quantum eigensolver (VQE) after extending a circuit, although circuit construction and action legality are deterministic and known. We introduce DreamQ...
69. Adaptive FastOPD: Progress-Aware Rollout Horizon Expansion for Efficient On-Policy Distillation ​
Author: Qian Tan, Huaifei Liang, Xuanyu Zhu, Lei Jiang, Yuqiang Li
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29494v1 Announce Type: new Abstract: On-policy distillation (OPD) provides dense teacher supervision along student-generated trajectories, but its online rollout process incurs substantial computational cost, particularly when a few long responses delay batch completion. Existing accelera...
70. Transcript-Managed Transformers: Monotone Multi-Agent Collapse and Universality with Two Pop-Enabled Transcripts ​
Author: Sergey Salishev
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.FL, cs.MA
arXiv:2607.29496v1 Announce Type: new Abstract: We study transcript management for fixed, finite-precision causal Transformers. A transcript is partitioned into channels of bounded blocks. Each transition consults a fixed visible suffix and may append one block, leaving the model, weights, and token...
71. The Grokked Illusion: True Equilibrium Mitigates Catastrophic Forgetting ​
Author: Xiaotian Zhang, Lai Shun Chan, Yue Shang, Entao Yang, Ge Zhang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29503v1 Announce Type: new Abstract: While neural networks are typically evaluated by their training and test performance, these metrics do not reveal how robust a learned representation is. Recent studies have shown that solutions occupying larger volumes in parameter space, as quantifie...
72. TerraNova: A Foundation Model for the Anthropocene ​
Author: Carlos Rodriguez-Pardo, Massimo Tavoni
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY, econ.EM, stat.ML
arXiv:2607.29527v1 Announce Type: new Abstract: A defining problem of the Anthropocene is to model the physical Earth and human societies as one coupled system, yet no learned representation spans their observational breadth. We argue the obstacle is geometric: the physical Earth is measured as cont...
73. A Neurosymbolic Approach for Explainable Early Diagnosis of Alzheimer's Disease ​
Author: Ranveer Singh, Pranuthi Tenali, Saurabh Mathur, Ameet Soni, Vaishali Phatak, Karla Lynch, Daniel Murman, Matthew Rizzo, Sriraam Natarajan
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29530v1 Announce Type: new Abstract: Identifying reliable Alzheimer's disease (AD) markers typically requires manual, labor-intensive transcription and expert analysis, limiting its scale. We introduce an automated pipeline that extracts qualitative knowledge about potential AD progressio...
74. Pyramidal Width Can Increase Under Vertex Insertion ​
Author: Jinze Zhao
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29555v1 Announce Type: new Abstract: Lacoste-Julien and Jaggi conjectured in 2015 that the pyramidal width of a polytope cannot increase when a vertex is added, provided that every old point remains a vertex. We give an exact counterexample with six integer points in $\R^3$. For [ P=\con...
75. MOT-SR: Multi-Objective Tool-Augmented Scientific Equation Discovery with Large Language Models ​
Author: Boxiao Wang, Runxiang Wang, Kai Li, Chongming Li, Zhiwei Chen, Yifan Zhang, Jian Cheng
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29561v1 Announce Type: new Abstract: Symbolic Regression (SR) aims to discover analytical equations from observational data and plays a central role in scientific modeling. While recent Large Language Model (LLM) based approaches show promise, they face two limitations. First, they lack d...
76. Convergence and Regret of the Policy Gradient for Multi-Armed Bandits in Diffusion Environment ​
Author: Yanwei Jia, Du Ouyang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2607.29593v1 Announce Type: new Abstract: This paper studies the policy gradient update for a multi-arm bandit problem in diffusion environment that is described by a stochastic differential equation (SDE) under the continuous-time reinforcement learning framework by Wang et al. (2020), Jia an...
77. The Parts Are Greater Than the Sum: Automated Task Sequencing for Efficient Training of Multi-Policy LLMs ​
Author: Jiajia Tang, Sizhe Yuen, Francisco Gomez Medina, Yali Du, Adam Sobey
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29601v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) commonly adapts large language models using a single shared Low-Rank Adapter (LoRA). This shared optimization space often suffers from interference when adapting heterogeneous task sequences, leading to poor trans...
78. A Human-Centered Validation of the Explainability-Performance Coefficient ​
Author: Christian Oliva, Luis F. Lago-Fern'andez
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2607.29614v1 Announce Type: new Abstract: The rapid adoption of deep learning models in high-risk domains has intensified the need for trustworthy Explainable Artificial Intelligence (XAI). However, objectively evaluating explanation fidelity and aligning XAI metrics with human-centered unders...
79. When Does On-Policy Interaction Help? Representational Tradeoffs in Value-Based Imitation Learning ​
Author: Luca Viano, Antoine Moulin, Audrey Huang, Volkan Cevher, Philip Amortila, Dylan J. Foster
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.29617v1 Announce Type: new Abstract: Imitation learning (IL)---training an agent to replicate expert behavior from demonstrations---underpins applications from robotics to language model training. Standard approaches such as Behavior Cloning (BC) are known to suffer from compounding error...
80. CENDRe: Concept Extraction with Natural Domain Representations ​
Author: Antonia Holzapfel, Andres Felipe Posada Moreno, Sebastian Trimpe
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.29621v1 Announce Type: new Abstract: Convolutional neural networks (CNNs) are widely used for time-series classification, but their deployment in critical domains requires understanding the temporal and spectral patterns that drive their predictions. Concept extraction (CE) methods identi...
81. GQ-FSL: Green Quantized Federated Split Learning ​
Author: Idan Roth, Lutz Lampe
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, eess.SP
arXiv:2607.29659v1 Announce Type: new Abstract: Deploying state-of-the-art deep neural networks (DNNs) at the wireless edge is severely bottlenecked by the strict energy and resource constraints of mobile devices. While federated split learning (FSL) mitigates on-device computation by offloading wor...
82. Freeze, Then Select: Structured Field Adapters and Stability-Validated Weak Selection for PDE Discovery from Sparse Observations ​
Author: Juncheng Zhong, Chenghuang Shen, Jianfeng Liu, Zhengdong Xiao, Longjiu Luo, Qianrong Wang, Wenjun Xu, Wenlian Lu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2607.29665v1 Announce Type: new Abstract: PDE discovery from sparse observations requires reconstructing a continuous field and selecting the correct differential terms. Our analysis of optimization paths in coupled neural PDE discovery reveals three behaviors: the exact support can persist to...
83. SVR: Self-Verifying Refinement via Joint Verdict-Confidence Reinforcement Learning for Adaptive Test-Time Compute ​
Author: Hongyu Chen, Liang Lin, Guangrun Wang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.28457v1 Announce Type: cross Abstract: Scaling test-time computation can improve language-model reasoning, but uniform budgets waste computation on easy inputs, while verifier-guided refinement relies on external feedback. We introduce Self-Verifying Refinement (SVR), an oracle-free multi...
84. Can LLMs Really Understand Item Difficulty Levels? Implications for Automated Item Generation Using LLMs ​
Author: Xinyi Wang, Hong Jiao, Ming Li, Sydney Peters, Hanna Choi, Tianyi Zhou, Qingshu Xu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.28634v1 Announce Type: cross Abstract: The estimation of item difficulty plays a key role in both formative assessment and large-scale high-stakes summative assessments. This study explores how large language models (LLMs) perform in predicting item difficulty levels using items from a la...
85. Imbalanced Data Clustering via Targeted Data Augmentation Using GMM and LLM ​
Author: Noor Khalal, Abdallah Alaa-Eddine Djamai, Imed Keraghel, Mohamed Nadif
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.28635v1 Announce Type: cross Abstract: In Natural Language Processing (NLP), dealing with underrepresented topics is challenging, especially in unsupervised tasks where clustering might not adequately capture minority topics. To tackle this challenge, our paper presents a novel unsupervis...
86. Learning Stateful Predictive Knowledge From Experience ​
Author: Yan Song, Xidong Feng, Bo Liu, Xinyu Cui, Haotian Fu, Zichen Liu, Mengyue Yang, Cheng Deng, Jian Zhao, Jun Wang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.28638v1 Announce Type: cross Abstract: As large language model (LLM) agents increasingly learn from experience, they primarily rely on trajectory-level reflection to extract insights. Viewed through the lens of predictive knowledge, we argue that this approach operates on episodic hindsig...
87. ThinkReset: Learnable Intermediate Interface Construction for Bounded-Context Long-Horizon Reasoning ​
Author: Fei Ding, Yongkang Zhang, Runhao Liu, Yuhao Liao, Zijian Zeng
Published: 8/3/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.28642v1 Announce Type: cross Abstract: Long chain-of-thought reasoning improves performance on complex problems, but it also introduces redundancy accumulation, context overflow, and error anchoring. We argue that under bounded context windows, the core bottleneck is not trajectory compre...
88. Geographically Weighted Surrogate Models for Rapid Small-Area Chronic Disease Estimation ​
Author: Aanya Gupta, Szandra P'eter, Sara Von Hoene, Emma Von Hoene, Taylor Anderson
Published: 8/3/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG, stat.AP
arXiv:2607.28655v1 Announce Type: cross Abstract: Small-area estimation (SAE) enables researchers and policymakers to identify spatial disparities in health outcomes, but survey-based SAE products carry an inherent lag. Gold-standard estimates such as CDC PLACES are released roughly two years after ...
89. Evaluating Federated Pre-Training: On the Reliability of Downstream Fine-Tuning and Intrinsic Evaluation ​
Author: Claudia Grosser, Maike Heuer, Denis Krompass, Thomas A. Runkler
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.28658v1 Announce Type: cross Abstract: Federated pre-training offers a way to train foundation models on private or distributed data without centralizing the underlying datasets. However, evaluating federated pre-training remains challenging because differences in client participation and...
90. NeuroSynth: A Biologically Inspired Continual Reinforcement Learning Architecture for Mitigating Catastrophic Forgetting ​
Author: Yash Kini
Published: 8/3/2026, 4:00:00 AM
Categories: cs.NE, cs.LG
arXiv:2607.28663v1 Announce Type: cross Abstract: Artificial Intelligence (AI) systems often perform well on isolated tasks but struggle under continual learning conditions, where training on new tasks can overwrite previously acquired knowledge, a failure mode known as catastrophic forgetting. Biol...
91. The Checking Problem: What must be true before AI ships in a regulated firm ​
Author: Prerit Ahuja
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.28666v1 Announce Type: cross Abstract: Enterprise AI programmes stall at a rate that is widely quoted and poorly explained. This paper measures the mechanism. Six document-heavy workflows of the kind performed daily in regulated financial services were run across four model families and t...
92. Fracture Risk Prediction in Adults Over 50 Years Old Using DXA and EHR: Comparison of Traditional and Machine Learning Models in Two Large Cohorts ​
Author: Jiahe Qian, Hao Dai, Kunyu Yu, Hexin Dong, Xing He, Erik A. Imel, Jiang Bian, Yifan Peng, Yi Liu
Published: 8/3/2026, 4:00:00 AM
Categories: stat.AP, cs.LG
arXiv:2607.28671v1 Announce Type: cross Abstract: Accurate fracture risk prediction is important for osteoporosis management, but commonly used clinical tools may not fully use information available in electronic health records (EHRs) and dual-energy X-ray absorptiometry (DXA) reports. We developed ...
93. How Hard Does It Think? Analyzing Step-Aware Reasoning Energy in LLM Chain-of-Thought Trajectories ​
Author: Hui Wei, Junda Wu, Sheldon Yu, Sizhe Zhou, Yizhu Jiao, Ming Zhong, Bowen Jin, Tong Yu, Shijia Pan, Jiawei Han, Julian McAuley
Published: 8/3/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.28674v1 Announce Type: cross Abstract: Understanding how computational effort is allocated across individual chain-of-thought (CoT) reasoning steps remains an open challenge: existing interpretability methods rely on output-level signals or collapse processing depth into a single trajecto...
94. TELLER: Dual-Path Iterative Preference Optimization for Table Entity Linking ​
Author: Yixin Peng, Kehao Li, Stefan Decker
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.28680v1 Announce Type: cross Abstract: Entity linking in tables matches short and ambiguous cell mentions to their corresponding knowledge-base entities. Existing approaches typically rely on data preprocessing pipelines that retain either compact or extensive table content as contextual ...
95. Accelerated Random-Sweep Gibbs Sampling for Gaussian Graphical Models via Dual Normal Factor Graphs ​
Author: Borna Khodabandeh, Mehdi Molkaraie
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT, stat.CO
arXiv:2607.28706v1 Announce Type: cross Abstract: We study the convergence properties of the random-sweep Gibbs sampler for Gaussian graphical models with a thin-membrane prior. We demonstrate that the convergence rate of the Gibbs sampler is significantly accelerated in the dual model, which is obt...
96. WaiT for the Signal: Simple Frequency-Aware Flow-Matching ​
Author: Krunoslav Lehman Pavasovic, Th'eophane Vallaeys, St'ephane Mallat, Giulio Biroli, Luke Zettlemoyer, Brian Karrer, Jakob Verbeek
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, stat.ML
arXiv:2607.28760v1 Announce Type: cross Abstract: As image generation models scale to ever higher resolutions, global coherence, local detail, and texture fidelity become critical axes for generation quality. However, standard flow matching treats all spatial frequencies uniformly, ignoring the natu...
97. Benchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation ​
Author: Philipp D. Siedler, Jordan Sassoon
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.28801v1 Announce Type: cross Abstract: Benchmark datasets are central to evaluating Large Language Models (LLMs), yet they are typically conceived as monolithic tasks, obscuring substantial variation in the demands of individual samples. We introduce a dataset-centric meta-evaluation fram...
98. Rolling With Resistance: Preference-Optimized LLM Counselors Can Trade Goal Persistence for Relational Attunement in Motivational Interviewing ​
Author: Weiying Chen, Junlong Shen, Zhexuan Tang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.28814v1 Announce Type: cross Abstract: In Motivational Interviewing (MI), a client's sustain talk (arguments for the status quo) calls for the counselor to roll with resistance, a move that can fail in two opposite ways: capitulation (abandoning the change agenda to preserve rapport) or c...
99. When Unlearning Fails: Reliable Data Deletion under Post-Training in Agent Networks ​
Author: Zihao Ding, Jun Huang, Liang Dong
Published: 8/3/2026, 4:00:00 AM
Categories: cs.NI, cs.LG
arXiv:2607.28829v1 Announce Type: cross Abstract: Self-improving federated agent networks keep training after deployment by collecting new trajectories with the current policy and feeding them back into later rounds. This closed loop makes unlearning harder than a one-time model repair. When a data ...
100. DeltaServe: Host-Agnostic Co-Serving of Inference and Fine-Tuning for LLMs ​
Author: Jiaxuan Chen, Jianshu She, Ye Yuan, Rajat Ghosh, Karan Gupta, Qirong Ho, Xue Liu, Oana Balmau
Published: 8/3/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2607.28848v1 Announce Type: cross Abstract: LLM serving systems are provisioned for peak load to meet strict latency targets, leaving substantial GPU compute idle whenever traffic falls below peak. We present DeltaServe, a host-agnostic co-serving design that converts this idle inference capac...
101. TextCloak: Thwarting Unauthorized LLM Exploitation via RL-Driven Unlearnable Text ​
Author: Chengshuai Zhao, Pingchuan Ma, Dawei Li, Bohan Jiang, Zhiyuan Yu, Zhen Tan, Huan Liu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CR, cs.LG
arXiv:2607.28862v1 Announce Type: cross Abstract: The rapid development of Large Language Models (LLMs) has led to significant advances across a wide range of language tasks, while simultaneously raising growing concerns about unauthorized data exploitation and privacy leakage. Unlearnable examples ...
102. Conditioning Tree-Based Diffusions and Flows for Probabilistic Tabular Regression ​
Author: Silas Koemen
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.28864v1 Announce Type: cross Abstract: Tree-based diffusion models fit flexible conditional predictive distributions for tabular regression without a neural density estimator, but they inherit their design defaults---noising path, parameterization, training distribution, features, sampler...
103. Learning to Predict Performance-induced Emotion Differences in Classical Piano Music ​
Author: Joann Ching, Gerhard Widmer
Published: 8/3/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, cs.MM
arXiv:2607.28876v1 Announce Type: cross Abstract: Music is often used as a medium for communicating emotion, with performers shaping perceived affect through interpretation. This study addresses the challenge of identifying and predicting subtle changes in perceived emotion that are exclusively due ...
104. Open-Source LLM-Driven Formal Verification: A Multi-Agent Pipeline for RTL Repair ​
Author: Ha Trung Tran
Published: 8/3/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, cs.SE
arXiv:2607.28877v1 Announce Type: cross Abstract: Verification consumes the majority of modern chip design effort, yet the formal verification tools that provide mathematical guarantees of correctness remain expensive and restrictively licensed. While large language models (LLMs) have shown promise ...
105. RareSense: Rarity-Aware Similarity Search for Anomaly Retrieval in Transactional Data ​
Author: Sidahmed Benabderrahmane, Talal Rahwan
Published: 8/3/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG
arXiv:2607.28879v1 Announce Type: cross Abstract: Similarity search over sparse set-valued data is often dominated by frequent background attributes because classical measures such as Jaccard, cosine, and Hamming compare objects through atomic overlap. IDF (Inverse document frequency) weighting part...
106. LayoutBench: Performance Benchmarking of Cloud Storage Layouts for Multimedia Data ​
Author: Debopam Sanyal, Hongjie Chen, Alexey Tumanov, Joshua Kimball
Published: 8/3/2026, 4:00:00 AM
Categories: cs.DC, cs.DB, cs.LG
arXiv:2607.28880v1 Announce Type: cross Abstract: Modern multimedia machine learning workloads increasingly store large-scale datasets in cloud object storage services such as AWS S3. How these samples are physically organized in storage (i.e.,storage layout) directly affects how quickly and cheaply...
107. To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing ​
Author: Amir M. Ebrahimi, Mohammed Mehedi Hasan, Aaditya Bhatia, Gopi Krishnan Rajbahadur, Ahmed E. Hassan
Published: 8/3/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2607.28887v1 Announce Type: cross Abstract: Large language models increasingly write and repair production code, yet evidence is mounting that their test-passing patches leave codebases harder to maintain. We identify one concrete source: deletion avoidance, the systematic tendency to retain c...
108. Structured Neural Chaos: An Adaptive Surrogate Modeling Framework for Functional Uncertainty Quantification and Global Sensitivity Analysis ​
Author: Isabel Corona Guevara, Yeping Hu
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.28903v1 Announce Type: cross Abstract: Variance-based global sensitivity analysis (GSA) plays a key role in uncertainty quantification by identifying the contributions of uncertain inputs to the variability of the model response. The repeated model evaluations required for these tasks are...
109. Visual Distribution Anchoring for Efficient Prompt Tuning ​
Author: Pouya Parsa, Raoof Zare Moayedi, Seongjin Choi
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.28967v1 Announce Type: cross Abstract: Prompt tuning adapts vision--language models with few trainable parameters, but existing approaches trade off efficiency and adaptation: static textual prompts can overfit source classes, image-conditioned prompts add per-instance computation, and mu...
110. Don't Contrast the Impossible: Region-Constrained Batching for Contrastive User Modeling on a Local Community Platform ​
Author: Seungho Han, Byeongchang Kim, Jin Yu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.28971v1 Announce Type: cross Abstract: Contrastive learning is widely used for user modeling in large-scale recommender systems, where standard in-batch negatives implicitly assume universal exposure that any user can be shown any item. On local community platforms such as Karrot, however...
111. Extrapolating the emergence of Hamiltonian chaos with random-feature Hamiltonian neural networks ​
Author: Jaesung Choi
Published: 8/3/2026, 4:00:00 AM
Categories: nlin.CD, cs.LG, physics.comp-ph
arXiv:2607.28977v1 Announce Type: cross Abstract: Machine learning of Hamiltonian dynamics has driven growing interest in Hamiltonian neural networks (HNNs), which encode Hamilton's equations of motion into the learning architecture. Despite this progress, it remains unknown whether such networks ca...
112. Point2Radio: A Foundation Model for Cross-Scene Radio Fields from Material-Aware Point Clouds ​
Author: Chaozheng Wen, Chenghong Bian, Hongze Chen, Jun Zhang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.NI, cs.AI, cs.CV, cs.LG
arXiv:2607.28994v1 Announce Type: cross Abstract: High-fidelity radio fields are typically simulated for every scene--transmitter configuration or fitted separately to each scene, failing to exploit propagation structures shared across environments. We present Point2Radio, a foundation model that le...
113. PaletteID: Prototype-Composed Semantic Identifiers for Multimodal CTR Prediction ​
Author: Huanyu Liu, Baining Chen, Hui Liu, Zengyang Li, Ziyi Huang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.29000v1 Announce Type: cross Abstract: Multimodal information can improve the accuracy of click-through rate (CTR) prediction and effectively alleviate item cold-start and long-tail problems. Recent studies commonly discretize pretrained multimodal embeddings into semantic identifiers (SI...
114. Persistent Convolution: A Topological Framework for AI Alignment Testing and Semantic Space Characterization ​
Author: Tyler Ashoff, Jordan Rodu
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.29008v1 Announce Type: cross Abstract: Modern opaque AI models prize performance over interpretability, which makes testing difficult. However, formal statistical tests conducted on a model's embedding space can provide robust characterizations of semantic structure, concept separation, a...
115. SciFigPlag-Bench: A Benchmark for Provenance-Aware Scientific Figure Plagiarism Detection ​
Author: Zhiying Cui, Minghao Yang, Linlin Gao, Jie Liu, Pengyuan Li
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.29124v1 Announce Type: cross Abstract: Scientific figures often encode the visual evidence behind scientific findings, yet figure plagiarism remains underexplored as a benchmarked multimodal evaluation problem. We present SciFigPlag-Bench, a benchmark for provenance-aware reasoning over s...
116. Have I Seen You? Embedding Behavior Signals Synthetic Face Dataset Membership ​
Author: Pawe{\l} Borsukiewicz, Daniele Lunghi, Wendk^uuni C. Ou'edraogo, Jacques Klein, Tegawend'e F. Bissyand'e
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.29144v1 Announce Type: cross Abstract: Synthetic face datasets are increasingly used to reduce privacy exposure and data access constraints in biometric recognition. Yet the generators that produce these datasets are trained on real faces, so synthetic data may still reveal their real sou...
117. Transpiler Autotuning with Predictive Models for Quantum Circuit Optimization ​
Author: Piotr Malkowski, Domenik Eichhorn, Joshua Ammermann, Rinor Kelmendi, Nick Poser, Patrick Hopf, Ina Schaefer
Published: 8/3/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, cs.SE
arXiv:2607.29145v1 Announce Type: cross Abstract: Quantum software engineering is an emerging research field focusing on efficiently embedding the quantum programming paradigm into existing software ecosystems. A key aspect of this field is the realization of quantum algorithms using gate-based prog...
118. Few-shot Deep Learning for Phase-Amplitude Aberration Correction in Transcranial Focused Ultrasound ​
Author: Minju Seol, Minjee Seo, Seonaeng Cho, Kyungho Yoon
Published: 8/3/2026, 4:00:00 AM
Categories: eess.IV, cs.LG
arXiv:2607.29182v1 Announce Type: cross Abstract: Transcranial focused ultrasound (tFUS) is a non-invasive technique that delivers focused acoustic energy through the skull for neuromodulation and therapeutic applications. However, the heterogeneous structure of the skull induces complex, patient-sp...
119. GALA: Generative Aligned Learning for Adaptive Multimodal Representation in the Taobao Shangou Recommender System ​
Author: Jiping Liu, Zhongmin Zhang, Zisen Sang, Zhijia Fang, Tao Ouyang, Ma Jiang, Shaopeng Liang, Zeyang Hou, Guodong Cao, Jia Jia
Published: 8/3/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.29213v1 Announce Type: cross Abstract: Modern recommender systems in food delivery increasingly leverage multimodal signals, including images, text, and user interaction histories, to enhance user experience, yet effective fusion of these heterogeneous modalities remains challenging, hind...
120. TAVI-TEC: An AI-Based Tool for Procedural Planning of Transcatheter Aortic Valve Implantation ​
Author: Alessandra Zerillo, Stefano Cannata, Diego Bellavia, Daniele Ciriello, Simone Manini, Salvatore Pasta, Caterina Gandolfo
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.29243v1 Announce Type: cross Abstract: Computed tomography angiography (CTA) is crucial for preprocedural TAVI planning, providing the anatomical information required for prosthesis sizing and vascular access assessment. As the volume of TAVI procedure increases, improving efficiency and ...
121. Simple-regret rates and minimax optimality of fixed-prior expected improvement in Mat\'ern and squared-exponential RKHSs ​
Author: Emmanuel Vazquez, S'ebastien Petit
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.NA, math.ST, stat.TH
arXiv:2607.29245v1 Announce Type: cross Abstract: We study the expected improvement (EI) policy for minimizing a deterministic objective function $f$ on a nonempty compact set $\mathcal X \subset\mathbb R^d$. We assume that $f$ belongs to the RKHS $\mathcal H_k$ of a continuous positive-semidefinite...
122. CalibratedRubric: Task-Adaptive Rubric Banks for Open-Ended LLM Evaluation ​
Author: Mengting Chen, Yanshu Sun, Wanting Liang, Beidi Luan, Rui Sun, Dezhi Chen, Jing Li, Zuo Bai
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.29252v1 Announce Type: cross Abstract: Reliable evaluation of open-ended LLM outputs requires fine-grained rubrics, yet expert curation is costly and difficult to scale. Existing automated pipelines rely on strict judge unanimity and binary variance filters, which cannot distinguish measu...
123. RTLCurator: Label-Efficient Data Curation for RTL Generation ​
Author: Siyang Cai, Cangyuan Li, Wenjing Chang, Kun Wang, Haoyu Gao, Yinhe Han, Ying Wang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2607.29283v1 Announce Type: cross Abstract: Training large language models (LLMs) to write register-transfer level (RTL) requires large corpora of paired specifications and code, and such data is scarce enough that most public corpora are now synthesized. Synthesis provides scale but not corre...
124. Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens ​
Author: Yi Luo, Rongzhi Gu, Jixun Yao
Published: 8/3/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.LG, cs.SD
arXiv:2607.29363v1 Announce Type: cross Abstract: Balancing sequence length, representational capacity, and long-horizon stability is a central problem in autoregressive (AR) speech and audio generation. Representations with higher frame rates or greater capacity can preserve more signal detail, but...
125. The Greedy Advantage in Finite-Horizon Bandits ​
Author: Kai Zhou, Michael Lingzhi Li, Kai Wang
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.29375v1 Announce Type: cross Abstract: Organizations increasingly rely on sequential experimentation to improve decision-making. While the multi-armed bandit literature has developed algorithms with strong asymptotic regret guarantees, many practical applications operate over finite and e...
126. PTP: Previous-Token Prediction based LLM Inversion for Near-Exact Prompt Reconstruction ​
Author: Pirzada Suhail, Nagasai Saketh Naidu, Atanu R Sinha, Amit Sethi
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.29378v1 Announce Type: cross Abstract: Large language models (LLMs) generate text by auto-regressively sampling the next token. This inherently leads to a many-to-many mapping between prompts and responses, complicating the task of inferring prompts from observed outputs. Prior work on LL...
127. Analytical and Bootstrap Confidence Intervals of Double Machine Learning: Simulation studies and an application to rural-urban difference in obesity prevalence ​
Author: Haozheng Xu, Siyuan Ma, Qingyan Xiang
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP
arXiv:2607.29456v1 Announce Type: cross Abstract: Double Machine Learning (DML) is a popular approach for treatment effect estimation in various settings, which allows a wide range of flexible machine learning methods to be used for nuisance parameter estimation while preserving valid inference. In ...
128. MoPET: Parameter-Efficient Mixture-of-Experts for Unified Medical Image Classification ​
Author: Sebastian Doerrich, Daniel W"urtinger, Francesco Di Salvo, Shyam Nandan Rai, Christian Ledig
Published: 8/3/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2607.29462v1 Announce Type: cross Abstract: Adapting deep learning models to profound clinical heterogeneity typically relies on parameter-efficient fine-tuning (PEFT) to avoid the severe overfitting associated with full end-to-end network updates. Although PEFT successfully navigates limited ...
129. Lightweight Neural Networks for Affordance Segmentation: Enhancement of the Decoder Module ​
Author: Simone Lugani, Edoardo Ragusa, Rodolfo Zunino, Paolo Gastaldo
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.PF
arXiv:2607.29473v1 Announce Type: cross Abstract: The deployment of deep neural networks for visual affordance segmentation on wearable robots poses may prove critical, due to some conflicting aspects of the problem. On one hand, affordance segmentation requires high-level abstraction capabilities, ...
130. Evidence-Type Competition: When Can Interventional Data Teach Language Models Causal Direction? ​
Author: Xining Xun
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.29484v1 Announce Type: cross Abstract: Interventional data is widely regarded as the gold standard for teaching models causal reasoning. We test this assumption in a fully controlled synthetic environment pitting observational correlation against causal effect, and find it fails instructi...
131. Leveraging Transfer Learning with Class-Specific Decoders for Laparoscopic Segmentation ​
Author: Priya Tomar, Aditya Parikh, Christian Bauckhage, Rafet Sifa
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.29509v1 Announce Type: cross Abstract: Effective multi-organ segmentation in surgical data requires learning the intricate anatomical features and alleviating the challenge of class imbalance, which results from relatively lower proportions of small and limitedly exposed structures. Recen...
132. Ordered-to-disordered transfer learning with graph neural networks for formation-energy and HOMO-LUMO gap prediction in high-entropy perovskite oxides ​
Author: Panupol Untarabut, Narjes Jomaa, Sylvian Cadars, Olivier Masson, Samuel Bernard, Assil Bouzid, Santanu Saha
Published: 8/3/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cond-mat.other, cs.LG
arXiv:2607.29510v1 Announce Type: cross Abstract: High-entropy perovskite oxides (HEPOs) represent a chemically complex class of materials with promising functional properties, yet their vast compositional space and, chemical/structural disorder pose significant challenge for accurate property predi...
133. DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat ​
Author: Ismayil Ismayilov, Atakan Kara, Kaan Oktay
Published: 8/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.29577v1 Announce Type: cross Abstract: Games and simulators make valuable benchmarks by turning decisions into measurable outcomes, but many current suites under-test rules-rich tactical reasoning: the ability to choose well when geometry, timing, resources, objectives, and rule interacti...
134. TOOD: Task-Aware Out-of-Distribution Score Calibration for Continual Learners ​
Author: Mostafa ElAraby, Samer B. Nashed, Liam Paull
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.29592v1 Announce Type: cross Abstract: The primary challenge of continual learning (CL) systems is to learn new tasks while remaining performant on previously learned tasks. A similarly important though less well-studied aspect of CL systems is their ability to distinguish inputs that are...
135. QASP: Query-Adaptive Robust Vector Search Policy ​
Author: Hakan Ferhatosmanoglu, Kushal Kumar, Tal Wagner, Andy Warfield
Published: 8/3/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.29606v1 Announce Type: cross Abstract: A fundamental challenge of vector search is achieving consistently high recall while minimizing computational costs. Fixed search parameters cause significant performance variance across queries, and conventional evaluation on average recall masks th...
136. Sign compression for Muon: SignMuon, MuonSign, and the Limits of Error Feedback ​
Author: Maria Smirnova, Alexey Kravatskiy
Published: 8/3/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2607.29674v1 Announce Type: cross Abstract: SignMuon compresses the Muon update to one bit per parameter by taking its elementwise sign, providing the most direct way to run a matrix-aware optimizer under an extremely low communication budget. It outperforms SignSGD in practice, yet it can asc...
137. Differentially Private Nonparametric Modal Learning with Applications to Regression and Clustering ​
Author: Arkajyoti Bhattacharjee, Arnab Auddy
Published: 8/3/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.ME, stat.ML, stat.TH
arXiv:2607.29675v1 Announce Type: cross Abstract: Density modes provide a localized and interpretable summary of multimodal distributions, but their estimation under rigorous differential privacy constraints remains largely unexplored. We study differentially private recovery of density modes for mu...
138. Tensor Data Scattering and the Impossibility of Slicing Theorem ​
Author: Wuming Pan
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2012.01982v3 Announce Type: replace Abstract: This paper proposes a standard way to represent sparse tensors. A broad theoretical framework for tensor data scattering methods used in various deep learning frameworks is established. This paper presents a theorem that is very important for perfo...
139. Beyond Black-Box Advice: Learning-Augmented Algorithms for MDPs with Q-Value Predictions ​
Author: Tongxin Li, Yiheng Lin, Shaolei Ren, Adam Wierman
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.PF
arXiv:2307.10524v3 Announce Type: replace Abstract: We study the tradeoff between consistency and robustness in the context of a single-trajectory time-varying Markov Decision Process (MDP) with untrusted machine-learned advice. Our work departs from the typical approach of treating advice as coming...
140. Tipping Point Forecasting in Non-Stationary Dynamics on Function Spaces ​
Author: Miguel Liu-Schiaffini, Clare E. Singer, Nikola Kovachki, Sze Chai Leung, Hyunji Jane Bae, Kamyar Azizzadenesheli, Anima Anandkumar
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, math.DS
arXiv:2308.08794v4 Announce Type: replace Abstract: Tipping points are abrupt, drastic, and often irreversible changes in the evolution of non-stationary and chaotic dynamical systems. For instance, increased greenhouse gas concentrations are predicted to lead to drastic decreases in low cloud cover...
141. Communication-Efficient Secure Aggregation in Decentralized Learning ​
Author: Sayan Biswas, Anne-Marie Kermarrec, Rafael Pires, Rishi Sharma, Milos Vujasinovic
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2405.07708v3 Announce Type: replace Abstract: Decentralized learning (DL) enables participants to collaboratively train models without a central server, yet it faces significant scalability challenges that demand sparsification to reduce the prohibitive communication costs of peer-to-peer exch...
142. On the Expressive Power of Sparse Geometric MPNNs ​
Author: Yonatan Sverdlov, Nadav Dym
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2407.02025v5 Announce Type: replace Abstract: Motivated by applications in chemistry and other sciences, we study the expressive power of message-passing neural networks for geometric graphs, whose node features correspond to 3-dimensional positions. Recent work has shown that such models can ...
143. Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations ​
Author: Yonatan Sverdlov, Ido Springer, Nadav Dym
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2410.06665v4 Announce Type: replace Abstract: This paper explores the characterization of equivariant linear layers for representations of permutations and related groups. Unlike traditional approaches, which address these problems using parameter-sharing, we consider an alternative methodolog...
144. Dual-Force: Enhanced Offline Diversity Maximization under Imitation Constraints ​
Author: Pavel Kolev, Marin Vlastelica, Georg Martius
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO
arXiv:2501.04426v2 Announce Type: replace Abstract: Offline diversity maximization under imitation constraints can transform demonstration data into a set of distinct behavioral policies, improving robustness to distribution shift without additional environment interaction. In practice, however, exi...
145. Dimensionality reduction for homological stability and global structure preservation ​
Author: Alexander Kolpakov, Igor Rivin
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MS
arXiv:2503.03156v4 Announce Type: replace Abstract: We propose DiRe, a force-directed dimensionality reduction framework designed to preserve global structure and homological features while remaining practical on modern hardware. The method combines an initial embedding with a graph-based layout opt...
146. Cooperative Variance Estimation and Bayesian Neural Networks for Disentangling Aleatoric and Epistemic Uncertainties ​
Author: Jiaxiang Yi, Miguel A. Bessa
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2505.02743v3 Announce Type: replace Abstract: Real-world data contains aleatoric uncertainty - irreducible noise arising from imperfect measurements or from incomplete knowledge about the data generation process. Mean-variance estimation networks can learn this type of uncertainty but require ...
147. StaQ: a Finite Memory Approach to Discrete Action Policy Mirror Descent ​
Author: Alex Davey, Alena Shilova, Brahim Driss, Riad Akrour
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2506.13862v2 Announce Type: replace Abstract: In Reinforcement Learning (RL), regularization with a Kullback-Leibler divergence that penalizes large deviations between successive policies has emerged as a popular tool both in theory and practice. This family of algorithms, often referred to as...
148. Towards White-Box Deep Wireless Sensing ​
Author: Xie Zhang, Yina Wang, Chenshu Wu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2507.21799v2 Announce Type: replace Abstract: The empirical success of deep learning has spurred its application to the radio-frequency (RF) domain, leading to significant advances in Deep Wireless Sensing (DWS). However, most existing DWS models remain black boxes, with ad-hoc architectures a...
149. Patch-Based 3D Variational Autoencoder for Super-Resolution of Turbulent Channel Flow ​
Author: Anuraj Maurya
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, physics.flu-dyn
arXiv:2507.22082v2 Announce Type: replace Abstract: Direct numerical simulation (DNS) accurately resolves all spatio-temporal scales of wall-bounded turbulence but becomes prohibitively expensive as the Reynolds number increases. Super-resolution (SR) provides a practical alternative by reconstructi...
150. Adaptive Policy Backbone via Shared Network ​
Author: Bumgeun Park, Donghwan Lee
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2509.22310v2 Announce Type: replace Abstract: Reinforcement learning (RL) has achieved impressive results across domains, yet learning an optimal policy typically requires extensive interaction data, limiting practical deployment. A common remedy is to leverage priors, such as pre-collected da...
151. A Hamiltonian driven Geometric Construction of Neural Networks via the Lognormal family, Application to Financial Fraud Detection and to Network Security ​
Author: Prosper Rosaire Mama Assandje, Landry Foka Marius, Arnaud Gires Fobasso Tchinda, Fr'ed'eric Barbaresco, St'ephane R. Gael Ekodeck, Serge Alain Ebele
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.25778v3 Announce Type: replace Abstract: We presents a method for constructing neural networks intrinsically on statistical manifolds via the lognormal distribution. We demonstrate this approach by formulating a neural network architecture directly on statistical manifold. The constructio...
152. Fisher Information, Training and Bias in Fourier Regression Models ​
Author: Lorenzo Pastori, Veronika Eyring, Mierk Schwabe
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, physics.data-an, quant-ph
arXiv:2510.06945v2 Announce Type: replace Abstract: Motivated by the growing interest in quantum machine learning, in particular quantum neural networks (QNNs), we study how recently introduced evaluation metrics based on the Fisher information matrix (FIM) are effective for predicting their trainin...
153. Reinforced sequential Monte Carlo for amortised sampling ​
Author: Sanghyeok Choi, Sarthak Mittal, V'ictor Elvira, Jinkyoo Park, Esmeralda S. Whitammer
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2510.11711v3 Announce Type: replace Abstract: This paper proposes a synergy of amortised and particle-based methods for sampling from distributions defined by unnormalised density functions. We state a connection between sequential Monte Carlo (SMC) and neural sequential samplers trained by ma...
154. In-situ Autoguidance: Eliciting Self-Correction in Diffusion Models ​
Author: Enhao Gu, Haolin Hou
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.17136v2 Announce Type: replace Abstract: The generation of high-quality, diverse, and prompt-aligned images is a central goal in image-generating diffusion models. The popular classifier-free guidance (CFG) approach improves quality and alignment at the cost of reduced variation, creating...
155. Monotone and Separable Set Functions: Characterizations and Neural Models ​
Author: Soutrik Sarangi, Yonatan Sverdlov, Nadav Dym, Abir De
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.23634v4 Announce Type: replace Abstract: Motivated by applications for set containment problems, we consider the following fundamental problem: can we design set-to-vector functions so that the natural partial order on sets is preserved, namely $S\subseteq T \text{ if and only if } F(S)\l...
156. A Novel XAI-Enhanced Quantum Adversarial Networks for Velocity Dispersion Modeling in MaNGA Galaxies ​
Author: Sathwik Narkedimilli, N V Saran Kumar, Aswath Babu H, Manjunath K Vanahalli, Manish M, Aik Beng Ng, Vinija Jain, Aman Chadha
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2510.24598v2 Announce Type: replace Abstract: Current quantum machine learning approaches often face challenges balancing predictive accuracy, robustness, and interpretability. To address this, we propose a novel quantum adversarial framework that integrates a hybrid quantum neural network (QN...
157. Dynamic Priors in Bayesian Optimization for Hyperparameter Optimization ​
Author: Lukas Fehring, Marcel Wever, Maximilian Splieth"over, Leona Hennig, Henning Wachsmuth, Marius Lindauer
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.02570v3 Announce Type: replace Abstract: Bayesian optimization (BO) is a widely used approach to hyperparameter optimization (HPO). However, most existing HPO methods only incorporate expert knowledge during initialization, limiting practitioners' ability to influence the optimization pro...
158. Robust Bidirectional Associative Memory via Regularization Inspired by the Subspace Rotation Algorithm ​
Author: Ci Lin, Tet Yeap, Iluju Kiringa
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2511.11902v2 Announce Type: replace Abstract: Bidirectional Associative Memory (BAM) trained with Bidirectional Backpropagation (B-BP) often suffers from poor robustness and high sensitivity to noise and adversarial attacks. To address these issues, we propose a novel gradient-free training al...
159. Optimal Resource Allocation for ML Model Training and Deployment under Concept Drift ​
Author: Hasan Burhan Beytur, Haris Vikalo, Kevin S Chan, Gustavo de Veciana
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.NI
arXiv:2512.12816v2 Announce Type: replace Abstract: We study how to allocate resources for training and deployment of machine learning (ML) models under concept drift and limited budgets. We consider a setting in which a model provider distributes trained models to multiple clients whose devices sup...
160. Symplectic Representation of Legendre Dynamics ​
Author: Robert Simon Fong, Gouhei Tanaka, Kazuyuki Aihara
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.19409v2 Announce Type: replace Abstract: Modern learning systems act on internal representations of data, yet how these representations encode underlying physical or statistical structure is often left implicit. In physics, symplecticity keeps Hamiltonian systems faithful to their phase-s...
161. Latent Sculpting for Zero-Shot Generalization: A Manifold Learning Approach to Out-of-Distribution Anomaly Detection ​
Author: Rajeeb Thapa Chhetri, Saurab Thapa, Avinash Kumar, Zhixiong Chen
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2512.22179v3 Announce Type: replace Abstract: Detecting previously unseen attacks remains a major challenge for machine learning-based intrusion detection systems. Deep models trained on network traffic often achieve high accuracy on known attacks but fail under distributional shift because th...
162. GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR ​
Author: Jiaying Zhang, Lei Shi, Jiguo Li, Jun Xu, Jiuchong Gao, Jinghua Hao, Renqing He
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2601.09361v4 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a key paradigm for improving large-scale reasoning models. Unlike supervised fine-tuning (SFT), RLVR exhibits distinct optimization dynamics and is sensitive to the preservation of pre-traine...
163. Low-Cost Hard-Label Adversarial Attack with Theoretical Foundations ​
Author: Jun Liu, Leo Yu Zhang, Fengpeng Li, Isao Echizen, Jiantao Zhou
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2601.14300v4 Announce Type: replace Abstract: Hard-label black-box attacks, relying solely on top-1 predictions, represent one of the most challenging yet practically threat models. Despite recent progress, existing approaches face two key limitations: (1) they overlook the critical role of in...
164. Matterhorn: Masked Time-to-First-Spike Encoding by Reassigning the Silent State for Sparse and Energy-Efficient Spiking Transformers ​
Author: Zhanglu Yan, Kaiwen Tang, Zixuan Zhu, Zhenyu Bai, Qianhui Liu, Yongxin Zhu, Weng-Fai Wong
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.22876v2 Announce Type: replace Abstract: Spiking neural networks (SNNs) promise energy-efficient inference for large language models (LLMs), yet most reported savings rely on compute-operation counts that overlook data movement. Energy characterization of representative spiking transforme...
165. Expert-Data Alignment Governs Generation Quality in Decentralized Diffusion Models ​
Author: Marcos Villagra, Bidhan Roy, Raihan Seraj, Zhiying Jiang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.02685v3 Announce Type: replace Abstract: Decentralized Diffusion Models (DDMs) route denoising through experts trained independently on disjoint data clusters, which can strongly disagree in their predictions. What governs the quality of generations in such systems? We present the first e...
166. GEMSS: A Variational Method for Discovering Multiple Sparse Solutions in Classification and Regression Problems ​
Author: Kate\v{r}ina Henclov'a, V'aclav \v{S}m'idl
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2602.08913v3 Announce Type: replace Abstract: In underdetermined regression and classification problems, multiple feature subsets often yield equivalent predictive performance. In applied settings, especially with $n \ll p$, high dimension or collinearities, it is valuable to provide a domain ...
167. Stem: Rethinking Causal Information Flow in Sparse Attention ​
Author: Lin Niu, Xin Luo, Linchuan Xie, Yifu Sun, Guanghua Yu, Jianchen Zhu, S Kevin Zhou
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.06274v2 Announce Type: replace Abstract: The quadratic computational complexity of self-attention remains a fundamental bottleneck for scaling Large Language Models (LLMs) to long contexts, particularly during the pre-filling phase. In this paper, we rethink the causal attention mechanism...
168. Wrong Code, Right Structure: Learning Netlist Representations from Imperfect LLM-Generated RTL ​
Author: Siyang Cai, Cangyuan Li, Haoyu Gao, Kun Wang, Yinhe Han, Ying Wang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR
arXiv:2603.09161v2 Announce Type: replace Abstract: Learning effective netlist representations is fundamentally constrained by the scarcity of labeled datasets, as real designs are protected by Intellectual Property (IP) and costly to annotate. Existing work therefore focuses on small-scale circuits...
169. LightningRL: Breaking the Accuracy-Parallelism Trade-off of Block-wise dLLMs via Reinforcement Learning ​
Author: Yanzhe Hu, Yijie Jin, Pengfei Liu, Kai Yu, Zhijie Deng
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.13319v2 Announce Type: replace Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising paradigm for parallel token generation, with block-wise variants garnering significant research interest. Despite their potential, existing dLLMs typically suffer from a rigid accu...
170. Maximum Entropy Behavior Exploration for Sim2Real Zero-Shot Reinforcement Learning ​
Author: Jiajun Hu, Nuria Armengol Urpi, Jin Cheng, Stelian Coros
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.25464v2 Announce Type: replace Abstract: Zero-shot reinforcement learning (RL) algorithms aim to learn a family of policies from a reward-free dataset, and recover optimal policies for any reward function directly at test time. Naturally, the quality of the pretraining dataset determines ...
171. From Physics to Surrogate Intelligence: A Unified Electro-Thermo-Optimization Framework for TSV Networks ​
Author: Mohamed Gharib, Leonid Popryho, Inna Partin-Vaisband
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AR
arXiv:2603.29268v3 Announce Type: replace Abstract: High-density through-substrate vias (TSVs) enable 2.5D/3D heterogeneous integration but introduce significant signal-integrity and thermal-reliability challenges due to electrical coupling, insertion loss, and self-heating. Conventional full-wave f...
172. EgoEverything: A Benchmark for Human Behavior Inspired Long Context Egocentric Video Understanding in AR Environment ​
Author: Qiance Tang, Ziqi Wang, Jieyu Lin, Ziyun Li, Barbara De Salvo, Sai Qian Zhang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.08342v2 Announce Type: replace Abstract: Long context egocentric video understanding has recently attracted significant research attention, with augmented reality (AR) highlighted as one of its most important application domains. Nevertheless, the task remains highly challenging due to th...
173. OpsLLM: Construction of Large Language Model for Software Operations with Multi-stage Learning ​
Author: Jingkai He, Pengfei Chen, Chenghui Wu, Shuang Liang, Ye Li, Gou Tan, Xidao Wen, Chuanfu Zhang, Fang Situ, Qi Zhou
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.02906v3 Announce Type: replace Abstract: In the field of software operations, Large Language Models (LLMs) have attracted increasing attention. However, existing research has not yet achieved efficient and effective endto-end intelligent operations due to low-quality data, fragmented know...
174. Predict-then-Diffuse: Adaptive Response Length for Compute-Budgeted Inference in Diffusion LLMs ​
Author: Michael Rottoli, Subhankar Roy, Stefano Paraboschi
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.04215v3 Announce Type: replace Abstract: Diffusion-based Large Language Models (D-LLMs) represent a promising frontier in generative AI, offering fully parallel token generation that can lead to significant throughput advantages and superior GPU utilization over the traditional autoregres...
175. A Nonlinear Singular Value Theory for Neural Networks ​
Author: Brian Charles Brown, Mauricio Munoz, Robert Bridges, David Grimsman, Sean Warnick
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.06938v2 Announce Type: replace Abstract: Recently Brown et al. [2025] established a singular value decomposition (SVD) for maps (especially nonlinear) satisfying certain norm conditions. We prove that most modern neural architectures admit this nonlinear SVD (NLSVD) representation---with ...
176. P-Flow: Proxy-gradient Flows for Linear Inverse Problems ​
Author: Zehua Jiang, Fenghao Zhu, Xinquan Wang, Chongwen Huang, Zhaoyang Zhang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2605.08328v3 Announce Type: replace Abstract: Generative models based on flow matching have emerged as a powerful paradigm for inverse problems, offering straighter trajectories and faster sampling compared to diffusion models. However, existing approaches often necessitate differentiating thr...
177. When Bits Break Recourse: Counterfactual-Faithful Quantization ​
Author: Chaymae Yahyati, Ismail Lamaakal, Khalid El Makkaoui, Ibrahim Ouahbi
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2605.17160v2 Announce Type: replace Abstract: Model quantization is widely used to reduce memory, latency, and deployment cost, and is typically judged by whether predictive accuracy is preserved. In decision systems that provide algorithmic recourse, however, accuracy preservation is not suff...
178. MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination ​
Author: Joss Armstrong
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.MA
arXiv:2605.22949v3 Announce Type: replace Abstract: Foundation-model pools are increasingly used as black-box responders in coordinated systems where a coordinator must decide which response to trust. Raw self-reported confidence is the natural signal, but is not comparable across models and becomes...
179. RAPNet: Accelerating Algebraic Multigrid with Learned Sparse Corrections ​
Author: Yali Fink, Ido Ben-Yair, Lars Ruthotto, Eran Treister
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.26854v2 Announce Type: replace Abstract: The scalable solution of large sparse linear systems is a bottleneck in scientific computing and graph analysis. While algebraic multigrid (AMG) offers optimal linear scaling, its performance is severely constrained by the trade-off between the spa...
180. Commit to the Bit: Reactive Reinforcement Learning Done Right ​
Author: Onno Eberhard, Claire Vernade, Michael Muehlebach
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.28276v2 Announce Type: replace Abstract: Reinforcement learning algorithms are commonly analyzed (and designed) under the Markov assumption. This is unrealistic, as most environments encountered in practice are either partially observable, or require function approximation that restricts ...
181. A Fully Convolutional Approach to Denoising 2D Correlation Spectra ​
Author: Nisar Nellikunnummel, Andi M Barbour, Lutz Wiegart, Tatiana Konstantinova, Anthony M DeGennaro
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2605.29975v2 Announce Type: replace Abstract: We present a fully convolutional denoising autoencoder (FC-DAE) tailored for two-dimensional representations of dynamic correlations that is applicable to many experimental techniques. Here, we demonstrate its performance on two-time intensity corr...
182. Multi-Scale Feature Attention Network for Polymer Classification Using Terahertz Spectroscopy ​
Author: Roshni Mahtani, Il'an Carretero, Daniel Moreno-Paris, Aldo Moreno-Oyervides, Laura Monroy, Oscar El'ias Bonilla-Manrique, Roc'io del Amor
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.06554v3 Announce Type: replace Abstract: Reliable polymer identification is essential for ensuring the quality and safety of recycled plastics, yet conventional sorting and spectroscopic techniques often struggle to deliver robust discrimination. Terahertz (THz) spectroscopy offers a prom...
183. Encoding the Euler Characteristic Transform ​
Author: Nello Blaser, Odin Hoff Gardaa, Lars M. Salbu, Elena Xinyi Wang, Bastian Rieck
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, math.AT
arXiv:2606.10824v2 Announce Type: replace Abstract: The Euler Characteristic Curve (ECC) records the Euler characteristic of a linearly embedded cell complex as a function of filtration height in a given direction, and the Euler Characteristic Transform (ECT) is the injective shape descriptor obtain...
184. APPO: Agentic Procedural Policy Optimization ​
Author: Xucong Wang, Ziyu Ma, Yong Wang, Yuxiang Ji, Shidong Yang, Guanhua Chen, Pengkun Wang, Xiangxiang Chu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.12384v2 Announce Type: replace Abstract: Recent advances in agentic Reinforcement Learning (RL) have substantially improved the multi-turn tool-use capabilities of large language model agents. However, most existing methods assign credit over coarse heuristic units, such as tool-call boun...
185. SPICE: Synergy and Partial Information Based Curriculum Evolution ​
Author: Ankush Pratap Singh, Houwei Cao, Yong Liu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.16639v2 Announce Type: replace Abstract: Multimodal learning exploits complementary information across heterogeneous modalities. The informativeness of each modality can vary widely across samples and training stages. Existing multimodal curriculum learning strategies often assume that th...
186. SqLinear: Balanced Square Partitioning Makes Linear Interaction Sufficient for Large-Scale Traffic Forecasting ​
Author: Yongfeng Su, Hongwen Li, Zijian Zhang, Ziquan Fang, Lu Chen, Christian S. Jensen, Hong Gao, Yinjun Han
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.21072v2 Announce Type: replace Abstract: Traffic prediction is a core task in intelligent transportation systems and urban-scale decision making. Despite the effectiveness of mainstream neural network-based methods, their deployment in real-world settings with thousands of traffic sensors...
187. DRIFT: Difficulty Routing Self-DIstillation with Rhythm-Gated Exploration and Success BuFfer Training ​
Author: Haoning Wang, Yiwei Liu, Haisen Luo, Dan Liu, Junxi Yin, Haotian Wang, Lei Zhang, Xiaoyu Tian, Shuaiting Chen, Yuansheng Song, Baoyan Guo, Xiongfei Yan, Bolan Yang, Chengwei Liu, Ming Cui, Jiong Chen
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.30345v2 Announce Type: replace Abstract: Enabling large language models to achieve stable self-improvement without external expert supervision remains a central challenge in complex reasoning tasks. Existing self-distillation and reinforcement learning methods lack explicit mechanisms for...
188. ECHO: Prune To Act, Trace To Learn With Selective Turn Memory In Agentic RL ​
Author: Zijun Xie, Binbin Zheng, Enlei Gong, Jihua Liu, Yuyang You, Lingfeng Liu, Jiayao Tang, Guanqun Zhao, Aoqi Hu, Zeyu Chen
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.31650v4 Announce Type: replace Abstract: Long-horizon language agents must repeatedly interact with tools, accumulate evidence, and make decisions under bounded context windows. Context-management methods make such rollouts feasible by simplifying past interactions through deletion, foldi...
189. Contravariance Theory: Strong Alignment for Minimal Solutions to Hard Tasks ​
Author: Dan Yamins, Aran Nayebi
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC
arXiv:2607.08561v2 Announce Type: replace Abstract: A series of results from the NeuroAI over the past fifteen years have raised core questions both about how to compare Deep Neural Network (DNN) models to the brain, and about how much convergent evolution to expect between artificial networks and r...
190. A Machine Learning Surrogate for Component Criticality Ranking in Interdependent Power-Communication Networks ​
Author: Sohini Roy, Xheni Hylviu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.08918v2 Announce Type: replace Abstract: Cyber-physical power systems are vulnerable to cascading failures caused by interdependencies between power and communication infrastructures. Because evaluating large N-k contingency sets with a high-fidelity simulator is computationally expensive...
191. Application of machine learning to monster level prediction in tabletop RPG game design ​
Author: Jolanta 'Sliwa, Jakub Adamczyk
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.09196v2 Announce Type: replace Abstract: Designing balanced adversaries is a central but labor-intensive task in tabletop role-playing game (TTRPG) development. In systems such as Pathfinder, each monster is described by many numerical attributes that jointly determine its power, summariz...
192. Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions ​
Author: David R. Wessels, Farhad Ramezanghorbani, David W. Romero, Alireza Moradzadeh, Olivia Viessmann, Maksim Zhdanov, John St. John, Ken Janik, David M Knigge, Yucheng Tang, Erik J Bekkers, Saee Gopal Paliwal
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, stat.ML
arXiv:2607.19378v2 Announce Type: replace Abstract: Subquadratic alternatives to attention require compromises when applied to multi-dimensional data: standard convolutions lack global receptive fields and input dependency, while recurrent models require rasterizing data such as images, volumes, and...
193. Self-Boosting Vision-Language Models with Noisy Student On-Policy Self-Distillation ​
Author: Shuai Wang, Daoan Zhang, Zhe Tang, Hao Cheng, Jiaheng Wei
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.23125v2 Announce Type: replace Abstract: Post-training enables vision-language models (VLMs) to understand human instructions and perform various downstream tasks. Current post-training methods usually rely on human-annotated data, distillation from external models, reinforcement learning...
194. AllocBench: Measuring Online Tool Allocation Capability in LLM Agents ​
Author: Daniel Wang, Andrew Xu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.23332v2 Announce Type: replace Abstract: Creating a reusable tool is an investment: an agent pays a fixed cost now in exchange for the potential of future reuse. Therefore, a user should prefer an agent that creates a small number of highly reusable tools, rather than many one-offs. We in...
195. WorldDiT: A Unified Diffusion Architecture for World and Action Modeling ​
Author: Sen Wang, R. Gnana Praveen, Bidhan Roy, Marcos Villagra
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2607.23909v2 Announce Type: replace Abstract: Many recent robot policies pursue stronger control by using large pretrained vision-language models (VLMs) as the action backbone. We introduce WorldDiT, a unified diffusion transformer architecture that couples action generation with visual world ...
196. Beyond Aggregate Risk: Role-Stratified Conformal Risk Control for LLM Tool Calls ​
Author: Md Ashikur Rahman, Md Arifur Rahman, Niamul Hassan Samin, Khandaker Rifah Tasnia, Md Hasibul Amin, Sifat Rahman Ahona, Juena Ahmed Noshin
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.24343v2 Announce Type: replace Abstract: Language-model agents act through structured tool calls whose arguments carry very different risks: untrusted content may legitimately shape an email body but should never set a recipient, account, command, or credential. Existing conformal risk co...
197. Conformal Cascade: Distribution-Free Accuracy Guarantees for Multi-Tier LLM Inference ​
Author: Yifan Dou, Shikan Lian, Shibo Li
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.25018v3 Announce Type: replace Abstract: Large language model (LLM) cascades reduce inference cost by routing easy queries to a small model and deferring hard queries to a larger one. Production cascades govern this deferral through a confidence threshold, but LLM confidence scores are mi...
198. A2TTA: Anchored-and-Agile Test-Time Adaptation for Evolving Traffic Sensor Networks ​
Author: Du Yin, Xiachong Lin, Yue Tan, Jinliang Deng, Estrid He, Hao Xue, Flora D. Salim
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.25875v2 Announce Type: replace Abstract: Traffic forecasting is important for efficient traffic management and route planning in smart cities. Existing traffic forecasting studies typically assume fixed sensor graphs, overlooking the continuous evolution of real-world traffic networks, e....
199. Temporally Centered SIGReg Improves Multi-Task LeWorldModel Learning: From Analysis to Method ​
Author: Chang Liu, Fei Suo, Yanzhou Jin, Yusuke Iwasawa, Yutaka Matsuo, Yaonan Zhu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2607.26924v2 Announce Type: replace Abstract: Recent work on LeWorldModel (LeWM) has shown that the Sketched Isotropic Gaussian Regularizer (SIGReg) enables stable end-to-end world-model learning from pixels by regularizing the latent marginal distribution toward an isotropic Gaussian, thereby...
200. What Can Latent World Models Know? Physical Parameter Identifiability in Multimodal Predictive Representations ​
Author: Kaizhen Tan, Xin Xu, Siru Tao, Yixiao Li, Hanzhe Hong, Yang Feng, Heqing Du
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2607.27017v3 Announce Type: replace Abstract: A central premise of latent world models is that predicting the future forces a representation to internalize the physics of its environment. Which physical quantities does a trained latent actually contain, and what decides this? We answer with co...
201. Rethinking EEG-Based Disease Diagnosis: Decoupling Instance Representation Learning from Subject-Level Supervision ​
Author: Zhiyuan Ma, Zeyuan Li, Zhiyi Lu, Jiacheng Hao, Youlang Du, Zhen Jiang, Xinche Zhang, Yuhao Sun, Xinke Shen, Sen Song
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.27274v2 Announce Type: replace Abstract: EEG-based disease diagnosis requires one prediction per subject, yet common pipelines segment recordings into short instances, inherit the subject label for every instance, and train instance-level classifiers. This assumes that all instances provi...
202. SE(3)-MeanFlow: Few-Step Protein Backbone Generation on Lie Groups ​
Author: Yikun Bai, Binghang Lu, Yikai Liu, Elaheh Akbari, Soheil Kolouri, Linxuan Wang, Ping He, Shuchan Wang, Ruqi Zhang, Guang Lin
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.27431v2 Announce Type: replace Abstract: Generative modeling of protein backbones promises the de novo design of proteins with prescribed structural and functional properties. Existing diffusion and flow-matching models produce high-quality backbones on SE(3)^N, but inference requires num...
203. RIPPLE: Generating Multi-Channel Phase, Not Recovering It ​
Author: Jaehyuk Lee, Yeajin Lee, Dayeon Shin, Donghun Lee
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.SD
arXiv:2607.27775v2 Announce Type: replace Abstract: Generative models synthesize magnitude spectra with high fidelity, while phase is delegated to a recovery module---Griffin--Lim, a vocoder, or a latent decoder---applied independently to each channel. For multi-channel waveforms this delegation is ...
204. TriShield: Zero-Utility-Loss Defense Against Privacy Backdoors in Federated Language Model Fine-Tuning via Orthogonal Gradient Projection and Optimizer State Entanglement ​
Author: Cheng Wei (Honor Device Co., Ltd., Shenzhen, China)
Published: 8/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.27940v2 Announce Type: replace Abstract: Federated fine-tuning of large language models (LLMs) enables collaborative training without exposing raw data. However, a recent attack, NeuroImprint, demonstrates that a malicious parameter server can corrupt a PEFT adapter into a privacy backdoo...
205. POSSE-kNN: Pathwise Out-of-Bag Selected Subspace Ensembles for Binary Classification ​
Author: Zardad Khan, Amjad Ali, Najd Adeed, Saeed Aldahmani
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2211.11278v3 Announce Type: replace-cross Abstract: Nearest neighbour classification is attractive for tabular data, but its performance can deteriorate when a fixed query centred neighbourhood does not follow the local class geometry. This study evaluates POSSE-$k$NN, a pathwise $k$ nearest n...
206. Information Processing by Neuron Populations in the Central Nervous System: A Theory of the Mathematical Structure of Data and Operations ​
Author: Martin N. P. Nilsson
Published: 8/3/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.LG, cs.NE
arXiv:2309.02332v3 Announce Type: replace-cross Abstract: In the mammalian central nervous system, neurons are organized into populations communicating by spike trains propagating along axonal bundles. How such populations encode and transform information is only partially understood. In this study ...
207. Nonparametric Partial Disentanglement via Mechanism Sparsity: Sparse Actions, Interventions and Sparse Temporal Dependencies ​
Author: S'ebastien Lachapelle, Pau Rodr'iguez L'opez, Yash Sharma, Katie Everett, R'emi Le Priol, Alexandre Lacoste, Simon Lacoste-Julien
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2401.04890v2 Announce Type: replace-cross Abstract: This work introduces a novel principle for disentanglement we call mechanism sparsity regularization, which applies when the latent factors of interest depend sparsely on observed auxiliary variables and/or past latent factors. We propose a r...
208. Unified continuous-time q-learning for mean-field game and mean-field control problems ​
Author: Xiaoli Wei, Xiang Yu, Fengyi Yuan
Published: 8/3/2026, 4:00:00 AM
Categories: math.OC, cs.LG, q-fin.CP
arXiv:2407.04521v3 Announce Type: replace-cross Abstract: This paper studies the continuous-time q-learning in mean-field jump-diffusion models in a setting where the environment simulator does not provide direct access to the population distribution. We propose the integrated q-function in decouple...
209. Deepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook ​
Author: Florinel-Alin Croitoru, Andrei-Iulian Hiji, Vlad Hondru, Nicolae Catalin Ristea, Paul Irofti, Marius Popescu, Cristian Rusu, Radu Tudor Ionescu, Fahad Shahbaz Khan, Mubarak Shah
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MM, cs.SD, eess.AS
arXiv:2411.19537v4 Announce Type: replace-cross Abstract: We survey deepfake generation and detection techniques, covering all deepfake media types: image, video, audio and multimodal content. We identify various kinds of deepfakes and construct taxonomies of deepfake generation and detection method...
210. Reproducing Human Individual Motor Signatures: A Data-Driven Approach for Repetitive Motion ​
Author: Angelo Di Porzio, Marco Coraggio
Published: 8/3/2026, 4:00:00 AM
Categories: cs.GR, cs.AI, cs.LG, cs.SY, eess.SY
arXiv:2503.15225v3 Announce Type: replace-cross Abstract: The deployment of autonomous virtual avatars (in extended reality) and robots in human group activities---such as rehabilitation therapy, sports, and manufacturing---is expected to increase as these technologies become more pervasive. Designi...
211. Artifact detection and localization in single-channel mobile EEG for sleep research using deep learning and attention mechanisms ​
Author: Khrystyna Semkiv, Jia Zhang, Maria Laura Ferster, Walter Karlen
Published: 8/3/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2504.08469v3 Announce Type: replace-cross Abstract: Current methods for detecting artifacts in sleep EEG range from threshold-based algorithms to machine learning approaches, yet applications remain limited for single-channel mobile EEG. We propose a convolutional neural network (CNN) model in...
212. Moment kernels: a simple and scalable approach for equivariance to rotations and reflections in deep convolutional networks ​
Author: Siqi Fang, Zachary Schlamowitz, Andrew Bennecke, Daniel J. Tward
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2505.21736v2 Announce Type: replace-cross Abstract: Translation equivariance is a central reason convolutional neural networks have been successful in computer vision. Other symmetries, such as rotations and reflections, are similarly important in fields such as biomedical image analysis, but ...
213. ASVSim (AirSim for Surface Vehicles): A High-Fidelity Simulation Framework for Autonomous Surface Vehicle Research ​
Author: Bavo Lesy, Siemen Herremans, Robin Kerstens, Jan Steckel, Walter Daems, Siegfried Mercelis, Ali Anwar
Published: 8/3/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2506.22174v3 Announce Type: replace-cross Abstract: The transport industry has recently shown significant interest in unmanned surface vehicles (USVs), specifically for port and inland waterway transport. These systems can improve operational efficiency and safety, which is especially relevant...
214. Fast Feature Field ($\text{F}^3$): A Predictive Representation of Events ​
Author: Richeek Das, Kostas Daniilidis, Pratik Chaudhari
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.RO
arXiv:2509.25146v2 Announce Type: replace-cross Abstract: This paper develops a mathematical argument and algorithms for building representations of data from event-based cameras, that we call Fast Feature Field ($\text{F}^3$). We learn this representation by predicting future events from past event...
215. Paris: A Decentralized Trained Open-Weight Diffusion Model ​
Author: Zhiying Jiang, Raihan Seraj, Marcos Villagra, Bidhan Roy
Published: 8/3/2026, 4:00:00 AM
Categories: cs.GR, cs.DC, cs.LG
arXiv:2510.03434v3 Announce Type: replace-cross Abstract: We present Paris, the first publicly released diffusion model pre-trained entirely through decentralized computation. Paris demonstrates that high-quality text-to-image generation can be achieved without centrally coordinated infrastructure. ...
216. Provable Diffusion Posterior Sampling for Bayesian Inversion ​
Author: Jinyuan Chang, Chenguang Duan, Yuling Jiao, Ruoxuan Li, Jerry Zhijian Yang, Cheng Yuan
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.NA, math.PR, math.ST, stat.TH
arXiv:2512.08022v2 Announce Type: replace-cross Abstract: We propose a novel diffusion-based posterior sampling method within a plug-and-play framework. Our approach constructs a probability transport from an easy-to-sample distribution to the target posterior via a diffusion process. To initialize ...
217. Embedding of Low-Dimensional Sensory Dynamics in Recurrent Networks: Implications for the Geometry of Neural Representation ​
Author: Vikas N. O'Reilly-Shah, Alessandro Maria Selvitella
Published: 8/3/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG
arXiv:2601.19019v3 Announce Type: replace-cross Abstract: Neural population activity in sensory cortex is organized on low-dimensional manifolds, but why such manifolds arise and what determines their geometry remain unclear. We model cortical populations as recurrent circuits driven by low-dimensio...
218. Incorporating data drift to perform survival analysis on credit risk ​
Author: Jianwei Peng (Humboldt-Universit"at zu Berlin), Stefan Lessmann (Humboldt-Universit"at zu Berlin, Bucharest University of Economic Studies)
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, q-fin.RM
arXiv:2601.20533v2 Announce Type: replace-cross Abstract: Survival analysis has become a standard approach for modelling time to default by time-varying covariates in credit risk. Unlike most existing methods that implicitly assume a stationary data-generating process, in practise, mortgage portfoli...
219. FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale ​
Author: Ajay Patel, Colin Raffel, Chris Callison-Burch
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2601.22146v3 Announce Type: replace-cross Abstract: Due to limited supervised training data, large language models (LLMs) are typically pre-trained via a self-supervised "predict the next word" objective on a vast amount of unstructured text data. To make the resulting model useful to users, i...
220. RAPiD: Reward-Guided Consistency Distillation of Diffusion Planners for Real-Time Autonomous Driving ​
Author: Ruturaj Reddy, Hrishav Bakul Barua, Junn Yong Loo, Thanh Thi Nguyen, Ganesh Krishnasamy
Published: 8/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.RO
arXiv:2602.07339v2 Announce Type: replace-cross Abstract: Diffusion-based trajectory planners can model multi-modal driving behavior, but their iterative denoising process introduces a latency bottleneck for real-time closed-loop deployment. We present RAPiD, a reward-guided consistency distillation...
221. Enabling Low-Latency Machine learning on Radiation-Hard FPGAs with hls4ml ​
Author: Katya Govorkova, Julian Garcia Pardinas, Vladimir Loncar, Victoria Nguyen, Sebastian Schmitt, Marco Pizzichemi, Loris Martinazzoli, Eluned Anne Smith
Published: 8/3/2026, 4:00:00 AM
Categories: hep-ex, cs.LG
arXiv:2602.15751v2 Announce Type: replace-cross Abstract: This paper presents an end-to-end demonstration of a viable, ultra-fast, radiation-hard machine learning (ML) application on FPGAs, which could be used in future high-energy physics experiments. We present a three-fold contribution, with the ...
222. Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization ​
Author: Theophilus Amaefuna, Hitesh Vaidya, Anshuman Chhabra, Ankur Mali
Published: 8/3/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, math.IT
arXiv:2603.00910v3 Announce Type: replace-cross Abstract: Layer-wise capacity in large language models is highly non-uniform: some layers contribute disproportionately to loss reduction, whereas others are nearly redundant. Existing layer-scoring methods provide sensitivity estimates but do not give...
223. OPERA: Online Data Pruning for Efficient Retrieval Model Adaptation ​
Author: Haoyang Fang, Shuai Zhang, Yifei Ma, Hengyi Wang, Cuixiong Hu, Katrin Kirchhoff, Bernie Wang, George Karypis
Published: 8/3/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG
arXiv:2603.17205v3 Announce Type: replace-cross Abstract: Domain-specific finetuning is essential for dense retrievers, yet not all data pairs contribute equally to the learning process. We introduce OPERA, a data pruning framework that exploits this heterogeneity to improve both the effectiveness a...
224. Estimating near-verbatim extraction risk in language models with decoding-constrained beam search ​
Author: A. Feder Cooper, Mark A. Lemley, Christopher De Sa, Lea Duesterwald, Allison Casasola, Jamie Hayes, Katherine Lee, Daniel E. Ho, Percy Liang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2603.24917v3 Announce Type: replace-cross Abstract: Recent work shows that standard greedy-decoding extraction methods for quantifying memorization in LLMs miss how extraction risk varies across sequences. Probabilistic extraction -- computing the probability of generating a target suffix give...
225. ActionParty: Multi-Subject Action Binding in Generative Video Games ​
Author: Alexander Pondaven, Ziyi Wu, Igor Gilitschenski, Philip Torr, Sergey Tulyakov, Fabio Pizzati, Aliaksandr Siarohin
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2604.02330v2 Announce Type: replace-cross Abstract: Recent advances in video diffusion have enabled the development of "world models" capable of simulating interactive environments. However, these models are largely restricted to single-agent settings, failing to control multiple agents simult...
226. TimeRFT: Stimulating Generalizable Time Series Forecasting for TSFMs via Reinforcement Finetuning ​
Author: Siyang Li, Yize Chen, Zijie Zhu, Yuxin Pan, Yan Guo, Ming Huang, Hui Xiong
Published: 8/3/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.CV, cs.LG
arXiv:2605.00015v2 Announce Type: replace-cross Abstract: Time Series Foundation Models (TSFMs) have demonstrated strong generalization capability and data efficiency in time series forecasting through large-scale pretraining. However, adapting TSFMs to downstream forecasting tasks remains challengi...
227. Leveraging Image Generators to Address Data Scarcity: The Gen4Regen Dataset for Forest Regeneration Mapping ​
Author: Gabriel Jeanson, David-Alexandre Duclos, William Larriv'ee-Hardy, No'e Cochet, Mat\v{e}j Boxan, Anthony Desch^enes, Fran\c{c}ois Pomerleau, Philippe Gigu`ere
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.RO
arXiv:2605.05627v2 Announce Type: replace-cross Abstract: Sustainable forest management relies on precise species composition mapping, yet traditional ground surveys are labour-intensive and geographically constrained. While Uncrewed Aerial Vehicles (UAVs) offer scalable data collection, the transit...
228. A Benchmark for Strategic Auditee Gaming Under Continuous Compliance Monitoring ​
Author: Florian A. D. Burnat, Brittany I. Davidson
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CY, cs.GT, cs.LG
arXiv:2605.06340v2 Announce Type: replace-cross Abstract: Continuous post-deployment compliance audits, mandated by emerging regulations such as the EU AI Act and Digital Services Act, create a class of strategic gaming distinct from the one-shot input/output gaming studied in prior work. Regulated ...
229. Quotient Semivalues for False-Name-Resistant Data Attribution ​
Author: Florian A. D. Burnat, Brittany I. Davidson
Published: 8/3/2026, 4:00:00 AM
Categories: cs.GT, cs.CR, cs.LG
arXiv:2605.07663v2 Announce Type: replace-cross Abstract: Data valuation methods allocate payments and audit training data's contribution to machine-learning pipelines; however, they often assume passive contributors. In reality, contributors can split datasets across pseudonymous identities, duplic...
230. Differentially Private Auditing Under Strategic Response ​
Author: Florian A. D. Burnat
Published: 8/3/2026, 4:00:00 AM
Categories: cs.GT, cs.CR, cs.LG
arXiv:2605.07674v2 Announce Type: replace-cross Abstract: Regulatory audits of AI systems increasingly rely on differential privacy (DP) to protect training data and model internals. We study audit design when the audited developer can strategically respond to the privacy-constrained audit interface...
231. Do LLMs Hold Their Values? MANTA: A Multi-Turn Adversarial Benchmark for Animal Welfare Reasoning ​
Author: Isabella Luong, Joyee Chen, Sankalpa Ghose, David Williams-King, Linh Le, Allen Lu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG
arXiv:2605.16301v3 Announce Type: replace-cross Abstract: Evaluating animal welfare reasoning in LLMs remains an open challenge despite rapid deployment in consumer and professional contexts where welfare considerations appear implicitly in everyday queries. Existing benchmarks such as AnimalHarmBen...
232. AI4BayesCode: From Natural Language Descriptions to Validated Modular Stateful Bayesian Samplers ​
Author: Jungang Zou, Alex Ziyu Jiang, Qixuan Chen
Published: 8/3/2026, 4:00:00 AM
Categories: stat.CO, cs.AI, cs.LG
arXiv:2605.18476v2 Announce Type: replace-cross Abstract: Coding and computation remain major bottlenecks in Markov chain Monte Carlo (MCMC) workflows, especially as modern sampling algorithms have become increasingly complex and existing probabilistic programming systems remain limited in model sup...
233. CompoSE: Compositional Synthesis and Editing of 3D Shapes via Part-Aware Control ​
Author: Habib Slim, Shariq Farooq Bhat, Mohamed Elhoseiny, Yifan Wang, Mike Roberts
Published: 8/3/2026, 4:00:00 AM
Categories: cs.GR, cs.LG
arXiv:2605.19350v2 Announce Type: replace-cross Abstract: Creating and editing high-quality 3D content remains a central challenge in computer graphics. We address this challenge by introducing CompoSE, a novel method for Compositional Synthesis and Editing of 3D shapes via part-aware control. Our m...
234. Statistical Inference for Stochastic Gradient Descent: Beyond Finite Variance ​
Author: Jose Blanchet, Peter Glynn, Wenhao Yang
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2605.26000v2 Announce Type: replace-cross Abstract: Stochastic gradient descent (SGD) is foundational to large-scale statistical learning and stochastic optimization. However, in some modern statistical learning problems, stochastic gradients can exhibit infinite-variance behavior. Consequentl...
235. LearnedCache: eBPF-Integrated Perceptron-Based Eviction Policies for the Linux Page Cache ​
Author: Zejia Qi
Published: 8/3/2026, 4:00:00 AM
Categories: cs.OS, cs.LG
arXiv:2605.26168v2 Announce Type: replace-cross Abstract: Any device that runs Linux uses the Linux page cache, a central pillar in OS and application performance, serving to reduce extraneous disk access. Many page cache eviction policies have been developed but remain bound by the rigidity of heur...
236. Physics from Video: Identifiability of Time-Invariant Second-Order ODEs under Minimal Trajectory Conditions ​
Author: Yuanyuan Wang, Wenjie Wang, Kun Zhang, Mingming Gong
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, stat.ML
arXiv:2606.00115v2 Announce Type: replace-cross Abstract: Bridging the gap between visual realism and physical understanding is a core challenge for video-based world models. We study the structural identifiability of continuous-time physical laws from raw pixels, focusing on whether an encoder-only...
237. Side-Channel Attacks Survive Noise Cancellation in 3D Printers ​
Author: Eric Yocam, Varghese Vaidyan, Micah Flack, Gurcan Comert, Judith L. Mwakalonge
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CR, cs.ET, cs.LG
arXiv:2606.13952v2 Announce Type: replace-cross Abstract: Active Motor Noise Cancellation (AMNC) is a noise-reduction feature shipped in commercial fused deposition modeling (FDM) 3D printers. Because it suppresses the acoustic emissions that side-channel attacks exploit, it has security-relevant si...
238. A Model-Driven Approach for Developing Families of Reinforcement Learning Environments ​
Author: Xiaoran Liu, Istvan David
Published: 8/3/2026, 4:00:00 AM
Categories: cs.SE, cs.LG
arXiv:2606.20324v2 Announce Type: replace-cross Abstract: Virtual training environments are software-intensive systems in which reinforcement learning (RL) agents learn, adapt, and demonstrate meaningful behavior. Virtual training environments offer a safe and cost-efficient alternative to training ...
239. The Metanym Game: A Self-Contained, Self-Consistent LLM Peer-Community Benchmark for Structural Intelligence ​
Author: David Nordfors
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2606.21008v2 Announce Type: replace-cross Abstract: The metanym game is a competitive word game for LLMs that measures structural intelligence against established cognitive-science constructs. No content is given in advance; the contestants create all of it -- a new kind of analogy test, analo...
240. EvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures ​
Author: Bu\u{g}ra Alperen Ulu{\i}rmak, Rifat Kurban
Published: 8/3/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.SE
arXiv:2606.30219v5 Announce Type: replace-cross Abstract: This paper presents a systematic survey and conceptual synthesis of the shared measurement problem underlying large language model (LLM) evaluation and AI safety: benchmark scores, reward signals, and safety metrics can improve while the capa...
241. 1-Lipschitz Neural Networks on Hadamard Manifolds ​
Author: Davide Murari, Marta Ghirardelli, Ben Adcock, Elena Celledoni, Brynjulf Owren, Carola-Bibiane Sch"onlieb
Published: 8/3/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA
arXiv:2607.19335v2 Announce Type: replace-cross Abstract: Controlling the Lipschitz constant of a neural network is a standard way to promote robustness and stability. Most existing constraining strategies are designed for Euclidean spaces. In this work, we construct and analyze a class of 1-Lipschi...
242. HijackKV: New Threat in Position-Independent KV Cache Reuse ​
Author: Yichi Zhang, Zhiqi Wang, Huan Zhang, Yuchen Yang
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.19957v2 Announce Type: replace-cross Abstract: Key-Value (KV) cache reduces inference latency in large language models (LLMs). Traditional prefix-based reuse has low cache hit rates across inference requests because it requires exact token and position matches. To improve efficiency, rece...
243. DualityCert: Verifier-Gated Language-Model Repair of Broken Duality Claims in Quantum Field Theory ​
Author: Xingyang Yu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG, hep-th
arXiv:2607.23614v2 Announce Type: replace-cross Abstract: We present DualityCert, a symbolic verifier for candidate Seiberg-duality claims in four-dimensional N=1 quiver gauge theories. The verifier evaluates 't Hooft anomaly matching, superpotential R-charge consistency, central-charge matching, an...
244. GNN-based Multi-Agent Control of Traffic Shockwaves in Sparse Vehicular Ad-hoc Networks ​
Author: Prachi Nandi, Madhuri Malakar, Sonakshi Satpathy, Pabitra Mohan Khilar
Published: 8/3/2026, 4:00:00 AM
Categories: cs.NI, cs.LG, cs.NA, math.NA
arXiv:2607.23792v2 Announce Type: replace-cross Abstract: Traffic shockwaves are stop-and-go waves that propagate upstream through the streams of vehicles and are one of the major causes of traffic congestion, fuel inefficiency, and increased accident rates in modern transportation systems. Although...
245. FORGE: Frame Orthogonality in Relevance Geometry for Long-Form Video Understanding ​
Author: Ghazal Kaviani, Ghassan AlRegib
Published: 8/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.25266v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have enabled long-form video understanding at a scale that was not previously possible. However, the density of relevant content decreases sharply as video sequence length increases, and exposing the m...
246. OneShot: Index-in-Ranking with Neural Scoring for Large-Scale Retrieval ​
Author: Ziwei Li, Shuyao Li, Xufeng Cai, Xue Zou, Yiming Ma, Huiting Lu, Wujie Yan, Zhichen Zhao, Yang Lu, Zhe Wang, Rui Luo, Zhengyu Su, Dan Zhang, Yimin Tan, Ji Liu
Published: 8/3/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.27475v2 Announce Type: replace-cross Abstract: In modern recommendation systems, retrieval serves as a primary stage responsible for filtering billions of candidate items down to thousands prior to refined ranking. To make this massive search effective and efficient, the system relies on ...
247. On a joint simultaneous learning of relevant feature subsets and subspaces in regression-like problems ​
Author: Illia Horenko
Published: 8/3/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2607.28080v2 Announce Type: replace-cross Abstract: We extend a recently introduced Entropy-Optimal Manifold Clustering (EOMC) to allow for a joint simultaneous identification of subsets and subspaces of relevant features in nonstationary and nonlinear regression problems. It is shown that the...