arXiv cs.LG - 2026-08-24 ​
195 items collected.
1. Bankruptcy Prediction via Hybrid Resampling and Stacking Ensemble Techniques with Explainable Artificial Intelligence (XAI)-Driven Analysis ​
Author: Obu-Amoah Ampomah, Edmund Fosu Agyemang, Kofi Acheampong, Louis Agyekum, Enock Adu Bonsu, Eric Nyarko
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, stat.AP
arXiv:2608.20343v1 Announce Type: new Abstract: This study develops and evaluates a bankruptcy prediction framework that integrates consensus-based feature selection, hybrid resampling, stacking ensembles, and explainable artificial intelligence to improve minority-class detection in severely imbala...
2. Machine Learning and ARIMA Model Averaging for Adaptive Public Health Forecasting: Comparative Evaluation and an Ontario COVID-19 Case Study ​
Author: Yushu Zou, Ye Li, Johra Moosa, Martin Grunnill, Samir N. Patel, Venkata R. Duvvuri
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, stat.AP
arXiv:2608.20406v1 Announce Type: new Abstract: Public health forecasts must respond to abrupt changes in surveillance data without over-extrapolating noise, reporting artifacts, or temporary trends. We evaluated autoregressive integrated moving average (ARIMA), random forest, and extreme gradient b...
3. From Thermal Preference Prediction to Adaptive Thermal Intervention: A Reinforcement Learning Approach Using Physiological and Environmental Sensing ​
Author: Isibor Kennedy Ihianle, Emmanuel Manu, Ehsan Asnaashari, Mojgan Jadidi, Pedro Machado, Amrit Sagoo, Ahmad Lotfi
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.20423v1 Announce Type: new Abstract: Personalised thermal comfort is essential for occupant wellbeing and for the development of more responsive building-control strategies, yet conventional Heating, Ventilation, and Air Conditioning (HVAC) systems rely on static setpoints and population-...
4. BF1: A Causal Dyadic Sparse-Attention Retrofit for Efficient Long-Context Transformers ​
Author: Hina Dixit
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.20427v1 Announce Type: new Abstract: Dense causal attention remains expensive at long context even when implemented with highly optimized exact kernels. We study BF1, a deterministic block-aligned dyadic sparse-attention route that combines a small exact local neighborhood, a global first...
5. Approximate Homomorphisms and Convergent Representations in Transducers ​
Author: Santiago Cifuentes
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.20428v1 Announce Type: new Abstract: We study the stability of minimal representations of controlled stochastic processes (in particular, transducers) under perturbations. This question is motivated by recent experiments finding predictive-state structure in the latent representations of ...
6. Wrong-Physics Backdoors in Neural PDE Operators ​
Author: Hanbing Liang, Fujun Liu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2608.20439v1 Announce Type: new Abstract: Neural PDE operators are increasingly trained on reusable solver archives, yet validation often relies on clean prediction error and parameter-agnostic plausibility checks. We introduce cross-parameter relinking, a data-poisoning primitive that makes a...
7. Decision Tree and K-Means Analysis of Raman Spectra for Edible Oils: A Physics-Informed AI Approach ​
Author: Amrita Shaw, Chandrasekar S. N., Sai Muthukumar V., Jhinuk Gupta, Deepak L. N. Kallepalli
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.20440v1 Announce Type: new Abstract: Authentication of edible oils in processed foods is important for food quality, fraud prevention, and regulatory compliance. This study establishes an integrated Raman spectroscopy and machine-learning framework that links intrinsic spectral organizati...
8. Shared Physics Responses Recover Hidden Rankings in Neural Operator Libraries ​
Author: Hanbing Liang, Fujun Liu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, physics.comp-ph
arXiv:2608.20441v1 Announce Type: new Abstract: Selecting the optimal neural-operator prediction during deployment is challenging when high-fidelity reference solutions are unavailable. We demonstrate that under a squared Hilbert-space loss, ranking a finite model library depends strictly on the low...
9. Stored in Optimizer State, Valued by Later Training: A Causal Account of Subliminal Trait Transfer ​
Author: Qinyang Xu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20442v1 Announce Type: new Abstract: Subliminal trait transfer allows a student model to acquire behavioral dispositions from teacher-generated data in which the trait is not semantically expressed. Recent work explains how such signals enter gradients, but not how they survive source rem...
10. Amortized Bandwidth Learning for Kernel Density Estimation under Logarithmic Score ​
Author: Junyi Liang, Hailiang Du
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20445v1 Announce Type: new Abstract: Kernel density estimation converts finite samples into probability densities, but its performance depends critically on bandwidth selection. Classical selectors prescribe the sample-to-bandwidth rule analytically or asymptotically, or solve a new optim...
11. Mutual information and sensitivity analysis for feature selection in customer targeting: a comparative study ​
Author: Nestor Barraza, Sergio Moro, Marcelo Ferreyra, Adolfo de la Pe~na
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, math.PR
arXiv:2608.20447v1 Announce Type: new Abstract: Feature selection is a highly relevant task in a data-driven knowledge discovery project. Several techniques have been developed aiming at finding the features that influence most an outcome to predict, including mutual information and, in recent years...
12. When Clean Data Hurts: Learning with Monotone Corruptions Beyond Binary Classification ​
Author: Julian Asilis, Shaddin Dughmi, Chirag Pabbaraju
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.20480v1 Announce Type: new Abstract: Optimal learners are tailored to exploit the i.i.d.\ data assumption underlying the classic PAC model. What if an i.i.d.\ training sample were corrupted with correctly labeled examples drawn from an otherwise unrelated, even adversarial source? This mo...
13. Metag: A dataset to build agentic meta-reviewing capabilities ​
Author: Anirudh Sundar, Min Chen, Divya Tadimeti, Gemma Zhang, Alice Li, Nigel Boachie Kumankumah, Pavan Uttej Ravva, Sadid Hasan, Somya Chatterjee, Pruthvi Prakash Navada, Xiao Wang, Yue Kang, Sulaiman Vesal, Larry Heck
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20488v1 Announce Type: new Abstract: AI tools increasingly support tasks across the scientific research cycle, from experiment design and manuscript preparation to peer review. At the same time, the continuing growth in conference submissions has increased the burden on meta-reviewers, wh...
14. Bern2Edge: A Neurosymbolic Compiler for Edge Deployment via Bernstein Polynomial Networks ​
Author: Malak Gamal El-Din, Yifan Zhang, Yasser Shoukry, Sitao Huang, Salma Elmalaki
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AR
arXiv:2608.20497v1 Announce Type: new Abstract: Deploying high-accuracy neural networks on resource-constrained edge devices remains challenging, as existing approaches treat training, compression, and hardware synthesis as separate stages, leaving a gap between software-trained models and efficient...
15. When Graph-JEPA Learns the Wrong Thing: Diagnosing and Repairing Category-Conditional Collapse ​
Author: Gollam Rabby, S"oren Auer
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20516v1 Announce Type: new Abstract: Joint-embedding predictive architectures are selected almost universally by linear probing and effective rank. We report a case where both read healthily while the representation carries zero usable instance information. We repair it, and a second fail...
16. Learning Exact NVIDIA SASS Encoders with $\mathbb{F}_2$ Linear Algebra ​
Author: Jiading Gai
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20532v1 Announce Type: new Abstract: NVIDIA provides a SASS disassembler but no public SASS assembler for recent data-center GPUs, limiting controlled machine-code rewriting. We present F2Asm, which learns exact 128-bit SASS encoders from paired disassembly and original CUBIN instruction ...
17. AgentDecarbonizer: Carbon-Aware Execution for AI Agents ​
Author: Leyi Yan, Shuangning Li, Sihang Liu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20566v1 Announce Type: new Abstract: AI agents extend large language models from single prompt-response interactions to long-running, goaldirected workflows that issue many model calls, invoke tools, and interact with external environments. These workflows enable tasks such as software re...
18. Faults That Fortify: CNN Adversarial Robustness via GPU Undervolting ​
Author: Behnam Omidi, Ahmad Tahmasivand, Husam Alsyouri, Saba Al-Sayouri, Chongzhou Fang, Ihsen Alouani, Khaled N. Khasawneh
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.CR
arXiv:2608.20572v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) face a dual challenge: vulnerability to adversarial attacks and prohibitive training cost. Adversarial training is effective but expensive, a burden that grows as learning shifts to the energy-constrained edge. This...
19. Provable Edge-of-Stability for Adam on a One-Dimensional Quadratic ​
Author: Yiman Fong, Heng Yang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC
arXiv:2608.20638v1 Announce Type: new Abstract: The edge-of-stability (EoS) phenomenon of Adam has been widely observed, while its underlying dynamical mechanism is not yet fully understood. We study uncorrected Adam on a one-dimensional quadratic, a clean setting where constant curvature isolates t...
20. Meta-clustering of milk mid-infrared spectra identifies dairy cow groups associated with negative energy balance in early lactation ​
Author: T. Touil, E. R. Paquet
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20653v1 Announce Type: new Abstract: Clustering methods have been used to identify distinct groups of milk samples, cows, or herds. Fourier-transform infrared (FTIR) spectroscopy, particularly mid-infrared (MIR) spectroscopy, has been applied to individual cow milk samples to predict vari...
21. RiskTraf: Risk-Extrapolated Residual Learning for Multi-Variate Traffic Flow Prediction ​
Author: Guangyu Wang, Zhidan Liu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.20656v1 Announce Type: new Abstract: Traffic sensors commonly record flow, speed, and occupancy, but standard traffic flow forecasting benchmarks and models rarely exploit all three raw measurements reliably. Although speed and occupancy provide sensor-native traffic-state information bey...
22. C-Score: Beyond Accuracy for Robustness Assessment in Semi-Supervised Learning under Open-World Unlabeled Contamination ​
Author: Tsao-Lun Chen, Chi-Cheng Fu, Han-Yi E. Chou, Shun-Feng Su
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.20667v1 Announce Type: new Abstract: Pseudo-label-based semi-supervised learning has achieved strong performance due to its simplicity and scalability. However, it is typically developed under a closed-world assumption that unlabeled data are drawn from the same distribution as labeled da...
23. Lightweight Adaptive ReduNet via Hyperspherical Manifold Learning ​
Author: Zhenglin Huang, Qifa Yan, Bin Dai, Xiaohu Tang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.20668v1 Announce Type: new Abstract: In recent years, a white-box neural network called ReduNet has been proposed, which employs the maximal coding rate reduction (MCR$^2$) principle to transform raw data into low-dimensional discriminative features via a forward layer-wise construction p...
24. Reinforcement Learning for Continuous-Time Jump Markov Decision Processes with Applications to Network Dynamic Pricing ​
Author: Huiling Meng, Ningyuan Chen, Xuefeng Gao
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20680v1 Announce Type: new Abstract: We study reinforcement learning (RL) in Continuous-Time Jump Markov Decision Processes (CTJMDPs) featuring general discrete state spaces (which need not possess a vector space structure) and continuous/discrete action spaces. The setup covers many well...
25. Geometric Regularization for Long-Tailed Semi-Supervised Learning via Gaussian Feature Bridges ​
Author: Hongyang He, Xinyuan Song, Yan Zhong, Daizong Liu, Yanbin Li, Yang-fan He, Wenqiao Zhang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20710v1 Announce Type: new Abstract: Real-world semi-supervised learning (SSL) often encounters significant challenges with long-tailed label distributions and noisy pseudo-labels, which hinder generalization and amplify confirmation bias. In this work, we introduce a novel framework, Gau...
26. Hidden Axis of Uncertainty: Latent-Posterior Alignment in Graph Neural Networks with Bayesian Output Layers ​
Author: Suk Hoon Choi, Damdae Park, Junhyuk Choi, Hyein Jung, Changsoo Kim, Ung Lee, Kyeongsu Kim
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20758v1 Announce Type: new Abstract: Bayesian Neural Networks (BNNs) with Bayesian output layers provide a principled and tractable framework for quantifying predictive uncertainty, yet the mechanisms shaping that uncertainty remain unclear. While conventional theory attributes uncertaint...
27. Fuzzy-MoE: Interpretable Regime-Conditioned Expert Routing for Non-Stationary Multivariate Time Series Forecasting ​
Author: Lan Guo, Jie Xiao, Zhao Su, Jun Shen, Haoran Li, Weixia Ma, Qingguo Zhou, Binbin Yong
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.20761v1 Announce Type: new Abstract: In non-stationary multivariate time series, different variables and samples often exhibit heterogeneous latent dynamic states, while existing deep forecasting models usually compress them into a unified end-to-end mapping, leading to suboptimal modelin...
28. Resolution-Consistent Greedy Neural Approximation on Infinite-Dimensional Spaces ​
Author: Pablo M. Bern'a, Antonio Falc'o, Diego Mond'ejar
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, math.FA
arXiv:2608.20812v1 Announce Type: new Abstract: We develop constructive approximation and learning guarantees for shallow neural models with infinite-dimensional inputs observed through finitely many coordinates. The analysis is based on a parameter-normalized neural dictionary and its associated we...
29. Scaling Muon for Diffusion Transformers ​
Author: Chenghao Li, Xiao Han, Xinxin Huang, Wei Liu, Boyang Li, Bing Xiao, Heran Zhang, Juanma Perez Rua, Ke Xu, Kangning Liu, Linjun Kuang, Na Li, Tan Wang, Tian Xie, Wei Peng, Yang Pei, Yifan Xu, Yuanhao Zhai, Yuwei Lin, Zhe Wang, Zihao He, Daniel Li, Junbiao Tang, Ziyang Jiang, Dake Chen
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.20818v1 Announce Type: new Abstract: The matrix-aware optimizer Muon improves large model training by balancing updates across singular directions, yet its scaling behavior and end-to-end efficiency on large Diffusion Transformers (DiTs) remain unclear. We first establish Muon's scaling b...
30. Nothing Changed but the Model: CellFill -- Bounded In-Cell Learning for Bit-Identical, Revocable Updates to Quantized LLMs ​
Author: Zifeng Liu, Zhiyong Du, Yaxin Lu, Yiming Mao, Zhenhe Wang, Wenqi Shi, Zhengkun Jing
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20873v1 Announce Type: new Abstract: Every way of teaching a deployed language model something new -- full fine-tuning, adapter merging, model editing -- replaces the released checkpoint, and with it every evaluation and cache that referred to those exact bits. We instead learn inside the...
31. Decoupling Policy Extraction for Offline Reinforcement Learning ​
Author: Xuyao Lin, Yixiang Shan, Jinru Duan, Tao Yang, Xinyu Zhao, Runyu Lei, Yiming Zhao, Jiaxin Fan, Zongbao Feng, Peng Jia
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2608.20909v1 Announce Type: new Abstract: Offline RL methods commonly jointly train the actor and critic, where the critic is used to guide the actor toward higher-value actions. This coupled learning process is well motivated in online RL, where an improved actor collects new data that can fu...
32. Training, learning and inference: unified dynamics of neural systems ​
Author: Mian Wang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20965v1 Announce Type: new Abstract: We define an atomic generation fact f=(u,tau,omega,z;rho), recording the origin, realized transformation, concrete occurrence, generated result and relation role. Compiled into a Generation-Fact Graph (GFG), these facts provide an AI-native, compilable...
33. A Critical Audit of Spatiotemporal Forecasting Benchmark Datasets and Baselines ​
Author: Kenneth Martin, Simon Heilig, Asja Fischer, Michel F. C. Haddad, Adam M. Sykulski, Moshe Eliasof
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.20980v1 Announce Type: new Abstract: Graph neural networks (GNNs) are routinely employed for short-range forecasting on multivariate time series with a spatial graph structure. Despite the availability of many alternative datasets, method innovations within this domain are predominantly a...
34. Jacobian-guided Noise Injection for Quantization Robustness in Large Language Models ​
Author: Deepanshu Pandey, Arnav Chavan, Nahush Lele, Sankalp Dayal, Deepak Gupta
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.20988v1 Announce Type: new Abstract: Quantization of Large Language Models (LLMs) is often hindered by the sensitivity of the self-attention mechanism to discretization errors. We identify the softmax operator as a bottleneck for quantization stability due to its sensitivity to outliers a...
35. Trojaning the Alignment: Stealthy Backdoor Attacks against Graph Foundation Models ​
Author: Minhua Lin, Zhicheng Gao, Yilong Wang, Hanqing Lu, Xiang Zhang, Suhang Wang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20991v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) on text-attributed graphs (TAGs) align graph representations with language semantics to support transferable graph learning. Despite these advantages, the backdoor vulnerability of GFMs on TAGs remains insufficiently unde...
36. Free-Probability Kernels for Zero-Rollout Hyperparameter Selection in Reservoir Computing ​
Author: Sara Malacarne, Andrea Ceni, Claudio Gallicchio
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, stat.ML
arXiv:2608.20998v1 Announce Type: new Abstract: Reservoir computing (RC) couples a fixed recurrent dynamical system with a trained lightweight readout, but this efficiency is partly lost during hyperparameter selection: the recurrent gain, input scale, and leakage rate determine the reservoir's stab...
37. RODE: A Radial-Orthogonal Decoupled Engine for Optimization ​
Author: Guoxiang Xu, Bince Qu, Qi Sun, Cheng Zhuo
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21024v1 Announce Type: new Abstract: Modern neural network training increasingly uses matrix-aware optimizers, yet their conditioned matrix step is typically added directly to the weight, jointly changing its norm and direction. This interaction matters because the current norm determines...
38. Designing a Robust LLM-Based Evaluation System for Agentic AI in Drug Discovery Through Human Alignment ​
Author: Emma Granqvist, Roc'io Mercado, Samuel Genheden
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21057v1 Announce Type: new Abstract: Agentic large language model (LLM) systems are reshaping scientific workflows in chemistry and drug discovery, but evaluating their open-ended, tool-augmented outputs remains a fundamental bottleneck. Reference-based metrics such as BLEU and ROUGE fail...
39. TracingFlow: A Simulation-Free Trajectory Inference Framework Based on Second-Order Dynamics ​
Author: Yuhao Sun, Zekun Wu, Zixun Huang, Peijie Zhou
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.GN
arXiv:2608.21070v1 Announce Type: new Abstract: Inferring continuous system evolution from sparse temporal snapshots is a key challenge in generative modeling and single-cell omics. While Optimal Transport (OT) is popular, existing frameworks are largely restricted to first-order dynamics, assuming ...
40. Causal Modeling of Adverse Pregnancy Outcomes via Adaptive LLM Proposals ​
Author: Kavimayil P. Komarasamy, Saurabh Mathur, Ameet Soni, David M. Haas, Kristian Kersting, Sriraam Natarajan
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21079v1 Announce Type: new Abstract: Adverse Pregnancy Outcomes (APOs) such as preterm birth and gestational diabetes can have long-term consequences for both the mother and child, yet an understanding of their causes remains elusive. Causal discovery in this domain is especially challeng...
41. FlatLand: Personalized Graph Federated Learning via Tailored Lorentz Space ​
Author: Jiahong Liu, Ram Samarth B B, Xinyu Fu, Menglin Yang, Weixi Zhang, Rex Ying, Irwin King
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21096v1 Announce Type: new Abstract: Federated learning enables privacy-preserving collaborative training, but highly heterogeneous client data remain challenging, especially in graph federated learning where clients possess structurally diverse graphs. Existing personalized federated lea...
42. BackDFL: A Unified Benchmark For Backdoor Attacks and Defenses In Decentralized Federated Learning ​
Author: Mouhamed Amine Bouchiha, Gregory Blanc, Yufei Han
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.DC
arXiv:2608.21137v1 Announce Type: new Abstract: Decentralized Federated Learning (DFL) promises trust-free collaborative learning by replacing the centralized parameter server with peer-to-peer model exchange. However, this architectural shift fundamentally reshapes the threat landscape. Without glo...
43. COEC: Calibrated Orthogonal-Equivalence Compensation for Structured Pruning of Large Language Models ​
Author: Peiqi Yu, Nam Ling, Wei Wang, Wei Jiang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21142v1 Announce Type: new Abstract: Structured pruning reduces the size and inference cost of large language models (LLMs) by removing weight columns, but the resulting output error can degrade accuracy. Existing training-free compensation methods use an additive bias or a single orthogo...
44. Capturing Cardiac Cyclicity through Phase-Equivariant Self-Supervised Learning ​
Author: Blaise Delaney, Dominic Dootson, Juan Jose Juan Castella, Salil Patel, Andrew Pfaff, Yuji Xing, Jonny Hancox, Karin Sevegnani
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21147v1 Announce Type: new Abstract: The cyclic structure of physiological processes offers a natural prior for self-supervised representation learning, and the cardiac cycle provides a particularly well-defined setting in which to exploit it. We derive a phase-equivariant self-supervised...
45. Thermo-FL: Thermal-Aware Robust Federated Fine-Tuning of Large Language Models for Edge AI ​
Author: Shiva Shrestha, Kazi Shaharair Sharif, Zongxing Xie, Jiajing Huang, Anhao Xiang, Honghui Xu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2608.21172v1 Announce Type: new Abstract: Federated fine-tuning enables large language models to adapt on edge devices without centralizing private data, but practical deployments must address hardware instability and adversarial update corruption together. Thermally constrained clients may th...
46. A Neurosymbolic Approach for Constructing Planning Domain Models from Clinical Narratives ​
Author: Ranveer Singh, Saurabh Mathur, Michael Skinner, Prasad Tadepalli, Kristian Kersting, Sriraam Natarajan
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21186v1 Announce Type: new Abstract: Surgical procedures such as laparoscopic appendectomy are complex, high-stakes processes, yet formalizing their workflows for decision support remains a significant challenge. Inducing probabilistic planning domain models in this setting is particularl...
47. Tydra: An Efficient Hybrid Model for Tabular Data ​
Author: Mieszko Komisarczyk, Saurabh Mathur, Maurice Kraus, Sriraam Natarajan, Kristian Kersting
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21199v1 Announce Type: new Abstract: Transformer-based tabular foundation models such as TabPFN achieve strong predictive performance but incur quadratic computational cost with context length. On the other hand, subquadratic SSM-based alternatives such as Hydra trade away accuracy for ef...
48. Curriculum-Aware Interpolate-then-Refine: Learned Physiological Time-Series Imputation under Realistic Missingness ​
Author: Yu-Chao Huang, Haochen Zhang, Nicholas Konz, Tianlong Chen
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.21207v1 Announce Type: new Abstract: Imputing physiological time series (arterial blood pressure, blood glucose, etc.) is essential for addressing the missingness that pervades clinical data. Yet modern imputation methods perform poorly in this domain: a recent benchmark found that simple...
49. TRACE-C: Rank-Calibrated Relational Anomaly Detection for Multi-Stream Operational Telemetry ​
Author: Matthew Faucher
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.21251v1 Announce Type: new Abstract: Operational telemetry can be jointly anomalous while every individual stream stays inside its familiar range. TRACE-C is an auditable strictly-prior rank-calibrated detector for aligned multi-stream telemetry: same-regime rolling median/MAD residuals f...
50. ConceptTS: LLM-Guided Concept Bottlenecks for Interpretable Multivariate Time-Series Forecasting ​
Author: Yichen Jiang, Yueqiao Chen, Dongyu Liu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21277v1 Announce Type: new Abstract: State-of-the-art multivariate time-series forecasters can model complex temporal and cross-variable dependencies, yet their opaque representations provide limited insight into why a particular forecast is produced. This lack of transparency restricts t...
51. SPARCL: Spectral Partitioned Analytic Continual Learning ​
Author: James Hartley, Zeropy Surio, Daniel Whitmore, Hannah Clarke, Thomas Reed
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21307v1 Announce Type: new Abstract: Analytic continual learning has emerged as a strong exemplar-free alternative to gradient-based class-incremental learning because it replaces iterative optimization with closed-form ridge updates. Yet the usual forgetting narrative, centered on stocha...
52. Rethinking Expressivity and Efficiency in Test-Time Training ​
Author: Zeyun Zhong, Joya Chen, Manuel Martin, Frederik Diederichs, Juergen Gall, Juergen Beyerer
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21308v1 Announce Type: new Abstract: Test-Time Training (TTT) enables long-context processing via continuous weight updates during inference, but current methods struggle to balance the expressivity of per-token update dynamics with the hardware efficiency of chunk-wise approximations. We...
53. Time-Aware Tranformer-Based Prediction Model for AECOPD ​
Author: Weihao Qu, Ling Zheng, Dongyang Wang, Jiacun Wang, Haowen Pan
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21324v1 Announce Type: new Abstract: The rapid symptom change of Acute exacerbation of chronic obstructive pulmonary disease (AECOPD) makes it critical to have time-sensitive prediction models. However, most current machine learning models studying AECOPD use clinical and laboratory data,...
54. Across-Design Uncertainty in Short Pricing Panels: Evidence from Simulated Price Trajectories ​
Author: Pedro Cadahia Delgado
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, econ.EM
arXiv:2608.21334v1 Announce Type: new Abstract: Short observational pricing panels can contain many observations while offering only a small number of distinct price movements. This paper studies the inferential consequences of that distinction in a synthetic data-generating process calibrated to a ...
55. Asymmetric Capacity Allocation in Self-Refinement Pipelines ​
Author: Zhuoyi Yang, Ian G. Harris, Salar Hashemitaheri, Cassie Huang, Yuangang Li, Hyunwoo Oh, Paul Dourish, Tony Givargis, Mohsen Imani, Li Zhang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21345v1 Announce Type: new Abstract: Self-refinement, typically structured as generation, critique, and revision, is a widely adopted paradigm for improving LLM generation and serves as a core mechanism in many LLM agents. While the three stages involve different cognitive demands, most e...
56. Exploratory As-Analyzed No-Detection of Culturally-Marked Predicate-Triggered PII Amplification in a Synthetic-English RAG Probe: A Predicate-Resource-Confounded Audit ​
Author: Yanhang Li, Zhichao Fan, Zexin Zhuang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.CY, cs.LG
arXiv:2608.20351v1 Announce Type: cross Abstract: We ask whether stereotype-loaded queries about culturally marked people leak more personal information from a retrieval-augmented generation (RAG) system than otherwise-equivalent neutral queries. We pre-register a four-culture audit (en-Anglo, es-LA...
57. NeuroStrata: An Electroencephalographic Connectivity-Aware Deep Representation Learning Framework for Dynamic Brain Network Analysis of Mental Stress ​
Author: Sayantan Acharya, Hamzeh Asgharnezhad, Abbas Khosravi, Douglas Creighton, Roohallah Alizadehsani, U Rajendra Acharya
Published: 8/24/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.LG
arXiv:2608.20354v1 Announce Type: cross Abstract: This study introduces NeuroStrata, a connectivity-aware deep representation learning framework for EEG-based mental stress analysis using Time-Varying Partial Directed Coherence (TV-PDC). Unlike conventional EEG classification approaches based on sta...
58. TriPLU: Bypassing the Gate with Direct Trilinear Product FFNs in Tiny Language Models ​
Author: He Zhang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.20360v1 Announce Type: cross Abstract: We study whether tiny decoder-only language models benefit from feed-forward layers that directly multiply learned feature projections. TriPLU, a Trilinear Product Linear Unit, replaces the usual gated FFN branch with a product-only degree-3 branch t...
59. Multilingual Verifier Bias in RLVR: Benchmark, Rollout Diagnosis, and the Cross-Lingual Selection Bottleneck ​
Author: Chenyu Zhou, Qiliang Jiang, Xu Zhou
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.20362v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a standard recipe for training large language models on mathematical reasoning, where an answer verifier serves as a language-neutral reward function. We show that this assumption fails in mult...
60. Harmonic Torsional Diffusion for Protein-Ligand Flexible Docking ​
Author: Maksim Zhdanov, Pavel Strashnov, Vladislav Kurenkov
Published: 8/24/2026, 4:00:00 AM
Categories: q-bio.BM, cs.LG
arXiv:2608.20366v1 Announce Type: cross Abstract: Molecular docking requires reasoning jointly about ligand pose and protein flexibility. Most diffusion-based docking models predict torsional updates with generic Euclidean heads that ignore the periodic geometry of angular variables. This mismatch i...
61. VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models ​
Author: Hyunwoo Kim
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.20374v1 Announce Type: cross Abstract: How precisely can we tell a language model how to feel? Most work on emotional generation answers with a discrete label - happy, angry, sad - which cannot express a target like "mildly downcast but calm." We instead specify the desired affect as a co...
62. TH-GNN: Heterogeneous Temporal Graph Neural Networks for LLM-Agent Shilling Attack Detection ​
Author: Shivam Swarup, Divya Prakash Shrivastava, Rakesh Thakur
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.20376v1 Announce Type: cross Abstract: LLM agents can now generate realistic shilling profiles, fluent reviews, and coherent ratings at scale, systematically defeating recommender-system defenses. Text-only detectors that flag semantic drift in review embeddings are blind to graph structu...
63. If It Walks Like an Arbitrage: Protocol-Agnostic Detection with Decidable Structural Equivalence ​
Author: Adam Khayam, Hamid Kolli, Mohamed Iguernalala, \c{C}agdas Bozman
Published: 8/24/2026, 4:00:00 AM
Categories: q-fin.CP, cs.CR, cs.LG
arXiv:2608.20377v1 Announce Type: cross Abstract: Ethereum transactions admit a canonical structural form. Each execution trace is built into an abstract syntax tree of token transfers grouped by call-frame nesting and reduced by a convergent term rewriting system of 15 rules to a unique canonical f...
64. Interpretable Information-Decomposed Brain Graph Learning for fMRI-based Disease Diagnosis ​
Author: Dengyi Zhao, Zhiheng Zhou, Zihan Wang, Guiying Yan, Xingqin Qi
Published: 8/24/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG
arXiv:2608.20380v1 Announce Type: cross Abstract: Resting-state functional magnetic resonance imaging (rs-fMRI) has enabled non-invasive mapping of functional brain interactions for computer-aided diagnosis, yet most existing approaches reduce inter-regional relationships to correlation-based edge w...
65. When Retrieval Fails Before It Begins: Structurally Indirect Prerequisite Eviction as a Retention Failure in Agentic Memory ​
Author: Minkyu Song
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.20400v1 Announce Type: cross Abstract: Agentic memory under a fixed budget involves two stages: retention and retrieval. Existing retrieval-centered paradigms implicitly assume necessary evidence survives eviction, but we challenge this by isolating a pre-retrieval failure mode: structura...
66. World models of environment, agent and joint agent-environment systems ​
Author: Manuel Baltieri, Filippo Torresan, Yivan Zhang, Alexander Boyd, Fernando E. Rosas
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.20401v1 Announce Type: cross Abstract: World models are a central component of model-based reinforcement learning. They are usually discussed in terms of what variables they predict, such as observations, rewards, states, latent or information states. We argue that there is a prior distin...
67. Robust Discovery of Coarse-Grained Continuum Equations from Microscopic Dynamics ​
Author: Partha Sarathi Mondal, Manav Kumar Jalan, Anish Kumar, Shradha Mishra
Published: 8/24/2026, 4:00:00 AM
Categories: cond-mat.soft, cond-mat.stat-mech, cs.LG
arXiv:2608.20404v1 Announce Type: cross Abstract: The discovery of governing partial differential equations (PDEs) directly from spatiotemporal data has emerged as a powerful tool for understanding the dynamics of complex systems. In this work, we apply PDE-SINDy to well-known phase-separating syste...
68. Rigorous Evaluation of Large Language Models for Malaria Drug Discovery: Trade-offs in Performance, Scale, and Resource Utility ​
Author: Marvellous O. Ajala (Magami Open Sciences Initiative), Zainab Ashimiyu-Abdusalam (Magami Open Sciences Initiative), Comfort Adesina (Magami Open Sciences Initiative)
Published: 8/24/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG
arXiv:2608.20418v1 Announce Type: cross Abstract: We introduce Malaria-Instruct, a curated instruction-following dataset derived from the ChEMBL Legacy Malaria corpus for Malaria virtual screening, and conduct a systematic evaluation of five open-source LLMs; Gemma-2 2B/9B, TxGemma-2B/9B, and LlaSMo...
69. Uncertainty propagation in auto-regressive random neural network models ​
Author: Janice Adams, Daniele Venturi
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NE, physics.comp-ph
arXiv:2608.20483v1 Announce Type: cross Abstract: We develop analytical and particle-based methods for uncertainty propagation in random neural network models, where both the inputs and network parameters are allowed to be random. Building on the piecewise-linear structure of the Leaky ReLU activati...
70. Keep Your Friends Close, and the Right Neighbours Closer: Disaster-Conditioned Kernel-Regularized Graph Attention for Building Damage Classification ​
Author: Fuad Hasan, Chul Min Yeum
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.20548v1 Announce Type: cross Abstract: Disaster damage is spatial: buildings rarely fail in isolation. Yet using spatial context for damage classification remains surprisingly underexplored, and many pipelines still rely primarily on per-building appearance cues even when the dominant unc...
71. aiXamine: Unified Black-Box Evaluation of Cross-Dimensional Trade-offs in LLM Safety, Security, and Privacy ​
Author: Fatih Deniz, Yazan Boshmaf, Dorde Popovic, Issa Khalil
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.20554v1 Announce Type: cross Abstract: The critical failure modes in deployed large language models (LLMs) are cross-dimensional: a model can score 99.3 in safety alignment while refusing one in three benign queries, or improve across every capability metric while losing 21 points in priv...
72. Learning Prostate Anatomy at Test Time for Cancer Detection in Micro-Ultrasound ​
Author: Obed Korshie Dzikunu, Mohammad Mahdi Abootorabi, Mohamed Harmanani, Paul F. R. Wilson, Emma Willis, Ferdinand Luger, Adam Kinnaird, Brian Wodlinger, Parvin Mousavi, Purang Abolmaesumi
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.20557v1 Announce Type: cross Abstract: Domain shift across clinical centers using different imaging hardware or acquisition protocols remains a fundamental barrier to deploying deep learning models for prostate cancer (PCa) detection. Existing test-time adaptation (TTA) methods address di...
73. Consistency Models for Fast MRI Reconstruction Using Regularization by Denoising ​
Author: Merve G"ulle, Junno Yun, Ya\c{s}ar Utku Al\c{c}alar, Mehmet Ak\c{c}akaya
Published: 8/24/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG, physics.med-ph
arXiv:2608.20561v1 Announce Type: cross Abstract: Diffusion models (DMs) have emerged as powerful generative priors for MRI reconstruction with promising results. Yet DM-based methods require extensive iterative refinement, limiting their practical deployment. Consistency models (CMs) provide a comp...
74. Conditional-Independence-Regularized Distributional Autoencoders for Mixed-Type Data ​
Author: Siyuan Tang, Gongjun Xu, Ji Zhu
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML
arXiv:2608.20562v1 Announce Type: cross Abstract: Mixed-type data containing both numerical and categorical variables arise in many scientific and real-world applications. Existing representation learning and generative modeling approaches typically focus either on reconstruction accuracy or uncondi...
75. FlavourBench: Ranking Frontier Language Models with Executable Culinary Ground Truth ​
Author: Josef Chen, Erim Hayretci
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG, cs.SE
arXiv:2608.20574v1 Announce Type: cross Abstract: Open-ended language-model benchmarks usually inherit a judge: a human preference panel, another model, or a brittle exact-match key. We introduce FlavourBench, an automated benchmark in which a versioned culinary system supplies dense, executable gro...
76. Keyed Provenance Watermarking with Complementary Lattice-Based Secure Aggregation for Federated Learning ​
Author: Xinyun Liu, Zhi Lu, Yu Chen, Ronghua Xu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.20580v1 Announce Type: cross Abstract: Federated learning (FL) is vulnerable to multi-level attacks. However, existing methods address them separately, leaving FL exposed to data leakage, unauthorized reuse, and malicious gradient manipulation. In this work, we propose an FL framework tha...
77. JuryProbe: An Empirical Consensus-Risk Diagnostic for Routing Reference-Free Factuality Judge Panels to Grounded Verification ​
Author: Tianxin Zhou, Ruixi Lin
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.20607v1 Announce Type: cross Abstract: Panels of inexpensive LLM judges increasingly make accept-or-escalate decisions. In factuality settings, accepting a claim because several reference-free judges agree can create a hidden risk: agreement may reflect shared false-negative blind spots r...
78. Dual-Cache Latent Space Communication between Heterogeneous Language Models ​
Author: Jiyao Liu, Qi Zhang, Yaoyi Jia, Ziwen Kan, Song Wang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.20617v1 Announce Type: cross Abstract: Multi-agent LLM systems split work across models, so answering often requires knowledge that sits in another agent's context: a Sharer has encoded information that a Receiver needs to complete its task. They usually communicate by exchanging text, wh...
79. Minimax Optimality of Score-Entropy Discrete Diffusion ​
Author: Cholyeon Cho, Yuchen Wu
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.20635v1 Announce Type: cross Abstract: Discrete diffusion models have demonstrated strong performance across a range of datasets, including natural language data and graph-structured data. Among many variants, score-entropy discrete diffusion (SEDD) has achieved particularly strong empiri...
80. MIL-BERT: Classification of Arbitrarily Large Text with Performance and Explanatory Guarantees ​
Author: John Cadigan, Dayne Freitag, Eric Yeh
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.20636v1 Announce Type: cross Abstract: Many text classification decisions are viable based on constituent excerpts alone. Taking inspiration from the field of multiple instance learning, we present an algorithm for training a neural network to classify text by selecting such excerpts. We ...
81. Predicting Resource Efficient Hamiltonian Decomposition for Continuous-Time Quantum Walk Simulations ​
Author: Mostafa Atallah, Rebekah Herrman, Zain H. Saleem
Published: 8/24/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.20660v1 Announce Type: cross Abstract: Simulating a continuous-time quantum walk (CTQW) on a graph in the circuit model of quantum computing requires decomposing its Hamiltonian into terms that can be Trotterized into hardware-native gates. We consider two such decompositions: the standar...
82. Amplifying the imaging power of digital sky surveys with space telescopes data and generative AI ​
Author: Sai Teja Erukude, Lior Shamir
Published: 8/24/2026, 4:00:00 AM
Categories: astro-ph.IM, astro-ph.GA, cs.AI, cs.LG
arXiv:2608.20666v1 Announce Type: cross Abstract: While Digital sky surveys provide excellent throughput of image data and can cover a large footprint, their imaging power is normally inferior to that of space-based telescopes. Space-based telescopes, on the other hand, provide excellent imaging pow...
83. Temporal Validity on Real Software Histories: Eliminating Stale-Fact Errors in Code-Assistant Memory over GitHub Fixes ​
Author: Neeraj Yadav
Published: 8/24/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CL, cs.LG
arXiv:2608.20685v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) has no model of time: when a fact changes across a coding session - a function is renamed, an endpoint moves, a dependency is bumped - RAG retrieves both the old and new value with near-identical similarity and ca...
84. CDRL: Certification-Driven Reinforcement Learning for Neutrino Flavor Model Discovery ​
Author: Piyush Jha, Jake Rudolph, Victoria Knapp-P'erez, Max Fieg, Aishik Ghosh, Vijay Ganesh
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO, hep-ph
arXiv:2608.20686v1 Announce Type: cross Abstract: Many scientific discovery problems require searching combinatorial hypothesis spaces under complex domain constraints. Reinforcement learning (RL) offers a promising approach, but existing methods rely on scalar rewards that provide limited informati...
85. PSK at WMT 2026 MIST: Task-Specialized QLoRA Adapters for Multilingual Summarization and Question Answering ​
Author: Srikar Kashyap Pulipaka
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.20757v1 Announce Type: cross Abstract: We describe the PSK submission to the WMT 2026 Multilingual Instruction Shared Task. Our system uses the 3.35B-parameter Tiny Aya Global model with three QLoRA adapters, one for each task. The adapters are trained on multilingual document-summary pai...
86. Rethinking Demonstration Unlearning in Imitation Learning for Robotics ​
Author: Jiazhuo Li, Yu Zhang, Yiming Fei, Kangkang Dong, Xiaojun Zhu, Houde Liu, Jinze Tao
Published: 8/24/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.20784v1 Announce Type: cross Abstract: Imitation learning for robotics depends on human demonstrations, some of which people may later ask to remove. Retraining without them is the natural reference, but its cost grows with policy and dataset scale, motivating cheaper operators that edit ...
87. CubicSplat: Differentiable Vector Graphics via Error-Bounded Forward Relaxation ​
Author: Chenglong Liu, Xin Zhang, Yimeng Zhu, Liyang He, Yixiao Ma, Yu Su, Zhenya Huang, Qi Liu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.GR, cs.CV, cs.LG
arXiv:2608.20803v1 Announce Type: cross Abstract: Vector graphics are prized for their resolution independence, compact storage, and direct editability, making differentiable optimization of their parametric primitives an attractive goal. Yet classical rasterization is discontinuous with respect to ...
88. Neuro-Geospatial Modelling of EEG Affective States Using Literature-Informed Environmental Context ​
Author: Utsav Poudel, Jagannath Aryal, Subramaniyaswamy Vairavasundaram
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.HC, cs.LG
arXiv:2608.20807v1 Announce Type: cross Abstract: Environmental exposures such as air pollution and greenness have been associated with affective and cognitive outcomes, but EEG and environmental datasets are rarely jointly georeferenced. We investigate whether literature-informed environmental prio...
89. Fine-tuning LLMs for Tourist Trajectory Prediction using Field Experiment Data ​
Author: Tatsuya Amano, Hirozumi Yamaguchi
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CY, cs.LG
arXiv:2608.20830v1 Announce Type: cross Abstract: Evaluating mobility interventions at tourist destinations requires predicting visitor behavior under varying conditions. Traditional methods struggle because tourist decisions depend heavily on context like weather and fatigue, yet models cannot gene...
90. SAC-Copula: Quality-Preserving Watermarking for Diffusion Language Models via Smooth Correlated Gumbel Fields ​
Author: Baixin Li, Haiyun He
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.CR, cs.LG
arXiv:2608.20839v1 Announce Type: cross Abstract: Watermarking diffusion language models (DLMs) requires mechanisms compatible with iterative parallel unmasking rather than autoregressive decoding. Existing sampling-based watermarking methods typically inject position-wise i.i.d. perturbations, whic...
91. Sharing the Control Authority Between Deep Reinforcement Learning and Model Predictive Control: Application to Multi-Class Transportation Networks ​
Author: Giray Onur, Azita Dabiri, Bart De Schutter
Published: 8/24/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2608.20858v1 Announce Type: cross Abstract: Transportation networks, in particular multi-class transportation networks (i.e., networks with mixed vehicle types), are complex systems that are challenging to control. Recently, Deep Reinforcement Learning (DRL), which learns control policies from...
92. ReCurveflow: A Flow Matching Framework that Learns Curved Reaction Trajectories to Predict Transition State Geometries ​
Author: Seungheun Baek, Mogan Gim, Jaewoo Kang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.20869v1 Announce Type: cross Abstract: Predicting transition states (TS) in chemical reactions is crucial, as they provide insights into reaction mechanisms. Recent work on TS prediction have focused on flow matching supervised on straight linear paths that do not align with actual reacti...
93. EviRank: Structured Relevance Evidence for Multimodal Image Re-ranking ​
Author: Enjun Du, Siyi Liu, Zirong Chen, Xinyu Zuo, Jinwen Luo, Ruiwen Tao, Lisheng Duan, Haijin Liang, Jin Ma, Junfu Pu, Yongqi Zhang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.20886v1 Announce Type: cross Abstract: Real-world image search queries are multimodal and compositional: ``find this shirt in pink'' specifies an entity to retain, an attribute to modify, and context to ignore. Yet existing re-rankers either compress such multifaceted relevance into an op...
94. Quantization-Aware Healing: A Practical Recipe for Recovering Compressed, 4-Bit LLMs ​
Author: Bakbergen Ryskulov, Iker Garc'ia-Ferrero, David Montero, David Jansen, Ali Hashemi, Jezabel R. Garcia, Antonio Tiene, Rom'an Or'us
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.PF
arXiv:2608.20953v1 Announce Type: cross Abstract: Serving large language models cheaply increasingly means shipping models that are both structurally compressed to a fraction of their parameters and quantized to 4 bits. Together these steps degrade reasoning, mathematics, coding, and long-context be...
95. TreeWY: Speculative Verification for Gated DeltaNet Hybrids ​
Author: Sneha Murthy Ghantasala
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.DC, cs.LG, cs.PF
arXiv:2608.20961v1 Announce Type: cross Abstract: Modern open models are hybrids: most layers are linear-attention (Gated DeltaNet, GDN) layers carrying a small fixed-size recurrent state instead of a growing key-value (KV) cache. This makes ordinary decoding memory-efficient, but hurts speculative ...
96. Training DeepFilterNet with Accurate Room Acoustic Simulations Improves Single-Channel Speech Enhancement ​
Author: Alessia Milo, Georg G"otz, Steinar Gu{\dh}j'onsson, Daniel Gert Nielsen, Jesper Pedersen, Finnur Pind
Published: 8/24/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, physics.comp-ph
arXiv:2608.20971v1 Announce Type: cross Abstract: We investigate how the realism of synthetic room impulse response (RIR) datasets affects the training of DeepFilterNet3 for single-channel speech enhancement. We compare a DNS4 image-source-method (ISM) RIR dataset with a higher-acoustic-fidelity dat...
97. From a Static Multi-Level Small Semantic Codebook to a Dynamic Single-Level Large Semantic Codebook for Generative Recommendation ​
Author: Tianlu Xie, Xin Ku, Mingjie Sun, Yunhao Sha, Lixiang Wang, Peng Wang, Yiyu Wang, Wenjin Wu, Zhaojie Liu, Peng Jiang, Wenwu Ou
Published: 8/24/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.21012v1 Announce Type: cross Abstract: Generative recommendation represents each item with a sequence of discrete Semantic IDs (SIDs) and predicts the sequence to retrieve the next item. Typical systems use multi-level residual quantization, which increases autoregressive decoding cost an...
98. COMET: Contrastive Motion-Enhanced Temporal Reasoning for Video Multimodal Large Language Models ​
Author: Chenghua Zhu, Zhaolu Kang, Qifan Shi, Siyan Wu, Kehan Jiang, Lei Wei, Lianyu Hu, Guangyuan Dong, Mingbo Yang, Rui Lu, Guibo Luo
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG
arXiv:2608.21030v1 Announce Type: cross Abstract: Video multimodal large language models have advanced significantly, yet fine-grained motion-temporal understanding remains fragile. The core bottleneck is not only sparse frame sampling, but also the lack of a complete temporal modeling pipeline for ...
99. AudioWorldSim: Realistic Binaural Audio Datasets For World Models ​
Author: Luis Vitor Zerkowski, Luiz Velho
Published: 8/24/2026, 4:00:00 AM
Categories: cs.SD, cs.LG
arXiv:2608.21075v1 Announce Type: cross Abstract: This technical report presents AudioWorldSim, an open-source platform designed to generate realistic binaural audio datasets and advance research in audio-based machine learning, particularly world models. Built as a custom extension of Meta's SoundS...
100. Llama-Mobile: Efficient 2.7-Bit Quantization of VLMs ​
Author: Luka Ribar, Jeevan Bhoot, Douglas Orr
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.21134v1 Announce Type: cross Abstract: Deploying vision-language models (VLMs) on mobile devices is challenging due to their significant memory and compute requirements. We present a framework for quantizing VLMs for efficient inference on resource-constrained hardware. Our approach combi...
101. Human-JEPA: A Human-Centric Vision Model that Perceives and Anticipates ​
Author: Hui Wei, Licai Sun, Guoying Zhao
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.21160v1 Announce Type: cross Abstract: Machines that understand humans should perceive the present and anticipate the future. Existing human-centric vision model are pretrained on human images, set the state of the art in static dense perception, so motion and anticipation are out of reac...
102. Beyond Imitation: Self-Improving Robot Policies via Off-Policy Q-Planning ​
Author: Varun Giridhar, Anant Khandelwal, Jeremy A. Collins, Ignat Georgiev, Animesh Garg
Published: 8/24/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.21204v1 Announce Type: cross Abstract: Behaviour Cloning (BC) has driven remarkable progress in robot manipulation, yet it is fundamentally limited by its inability to self-improve: a policy that fails cannot learn from that failure without additional human demonstrations. Reinforcement L...
103. No PUN Intended: Plausible Unknown Names for Person-Centred LLM Evaluation ​
Author: Dimitri Staufer, David Hartmann, Ibrahim Baroud
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.21206v1 Announce Type: cross Abstract: Person names are widely used as prompt variables in LLM evaluations of factuality, privacy leakage, bias and abstention, but when a name's evidential status is uncontrolled, measurements may conflate memorisation, retrieval, name priors and wrong-per...
104. Personalized Privacy Control in LLMs via Attention Head Intervention ​
Author: Junseok Kim, Nakyeong Yang, Kyomin Jung
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.21209v1 Announce Type: cross Abstract: The rise of agentic AI enables LLMs to access diverse user data, raising critical privacy concerns. Prior work on contextual privacy studies whether LLMs regulate information disclosure according to context-dependent norms. However, acceptable disclo...
105. Event-triggered Implicit Perturbation for Zeroth-Order Fine-Tuning of Spiking Transformers ​
Author: Tengteng Lei, Prabodh Katti, Rashi Dutt, Houssem Sifaou, Tan Peng, Osvaldo Simeone, Kai Xu, Bipin Rajendran
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, cs.NE
arXiv:2608.21223v1 Announce Type: cross Abstract: Zeroth-order (ZO) optimization estimates gradients using only forward-pass evaluations, making it suitable for fine-tuning non-differentiable, event-driven spiking neural networks (SNNs). However, its deployment on in-memory computing (IMC) accelerat...
106. Advanced Linear Algebra with Applications - Part I (Numerical linear algebra for PDEs, machine learning, and data assimilation) ​
Author: Victorita Dolean, Jemima Tabeart
Published: 8/24/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA
arXiv:2608.21234v1 Announce Type: cross Abstract: These lecture notes form the first part of a master's-level course on advanced numerical linear algebra. Their aim is not only to present the classical algorithms, but to show why the subject has become considerably more central than it was a generat...
107. On the Transferability of Agricultural Weed Detection Under Cross-Field Distribution Shift ​
Author: Nikhilesh Prabhakar, Pranuthi Tenali, Wilfredo Abudeye Fernandez, Shekhar Borah, Athresh Karanam, Erik Blasch, Prabha Sundaravadivel, Sriraam Natarajan
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.21254v1 Announce Type: cross Abstract: Accurate agricultural weed detection in real-world field conditions is essential for precision agriculture, enabling targeted intervention and reducing yield loss. Recent work has reported strong detection performance from UAV-based imagery across a ...
108. The Exceedance Design Effect: Effective Sample Size for Thresholds under Clustering ​
Author: Adam Noonan
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.21262v1 Announce Type: cross Abstract: Many machine-learning systems set a threshold at a quantile of a calibration set: conformal predictors that promise 90% coverage by drawing their cutoff at the calibration set's 90th percentile, abstention gates that decline to answer when a model's ...
109. TurboBias 2.0: Streaming Context-Biasing for Production-Efficient ASR Systems ​
Author: Vladimir Bataev, Lilit Grigoryan, Andrei Andrusenko, Nikolay Karpov, Vitaly Lavrukhin, Boris Ginsburg
Published: 8/24/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.CL, cs.LG, cs.SD
arXiv:2608.21343v1 Announce Type: cross Abstract: Contextualization is essential for production automatic speech recognition (ASR) systems, where user-provided phrases must be recognized accurately under strict latency constraints. Although many context-biasing methods improve recognition accuracy, ...
110. Truthful Calibration Measures for Sequential Prediction ​
Author: Anagha Gokul, Jason Hartline, Lunjia Hu, Jonathan Ullman, Yifan Wu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.DS, cs.GT, cs.LG
arXiv:2608.21348v1 Announce Type: cross Abstract: Calibration requires probabilistic reports to be conditionally unbiased and reliably interpretable as probabilities. A calibration measure assigns numerical error to miscalibrated reports. Haghtalab et al. (2024) proposed an approximately truthful ca...
111. PerturbRx: Learning Treatment-Conditioned Latent Transitions for Patient Drug Response Prediction ​
Author: Yoshitaka Inoue, Minoh Jeong, Alfred Hero, Rui Kuang, Augustin Luna
Published: 8/24/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG
arXiv:2608.21349v1 Announce Type: cross Abstract: Scarce data and tumor heterogeneity limit patient-level cancer treatment-response prediction. Existing approaches predict response from pretreatment molecular profiles and drug representations, without explicitly modeling the molecular changes expect...
112. Primal Acceleration of Newton's Method ​
Author: Nikita Doikov
Published: 8/24/2026, 4:00:00 AM
Categories: math.OC, cs.AI, cs.LG
arXiv:2608.21359v1 Announce Type: cross Abstract: We develop a new direct accelerated Newton method for minimizing convex functions with Lipschitz continuous Hessian. The algorithm uses only primal variables and performs just one linear solve per iteration. With a simple predetermined choice of para...
113. Exact and general decoupled solutions of the LMC Multitask Gaussian Process model ​
Author: Olivier Truffinet (CEA Saclay), Karim Ammar (CEA Saclay), Jean-Philippe Argaud (EDF R&D), Bertrand Bouriquet (EDF)
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2310.12032v4 Announce Type: replace Abstract: The Linear Model of Co-regionalization (LMC) is a very general multitask gaussian process model for regression or classification. While its expressiveness and conceptual simplicity are appealing, naive implementations have cubic complexity in the p...
114. Forecasting with an N-dimensional Langevin Equation and a Neural-Ordinary Differential Equation ​
Author: Antonio Malpica-Morales, Miguel A. Dur'an-Olivencia, Serafim Kalliadasis
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, math.DS, physics.data-an, stat.ME
arXiv:2405.07359v2 Announce Type: replace Abstract: Accurate prediction of electricity day-ahead prices is essential in competitive electricity markets. Although stationary electricity-price forecasting techniques have received considerable attention, research on non-stationary methods is comparativ...
115. Federated and differentially private estimation of KL divergence ​
Author: Sayan Biswas, Graham Cormode, Carsten Maple, Mary Scott
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.DB
arXiv:2411.16478v3 Announce Type: replace Abstract: Measuring distribution drifts is a key task in managing distributed, sensitive data, as it underpins a wide range of federated learning and analytics applications. In many practical settings, however, directly sharing such information is either und...
116. Structure is information: structural identifiability mappings for machine learning with partially observed dynamical systems ​
Author: Janis Norden, Elisa Oostwal, Michael Chappell, Peter Tino, Kerstin Bunte
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2502.04131v2 Announce Type: replace Abstract: The successful application of modern machine learning for time series classification is often hampered by limitations in quality and quantity of available training data. To overcome these limitations, domain knowledge can be leveraged in the form o...
117. Smart Exploration in Reinforcement Learning using Bounded Uncertainty Models ​
Author: J. S. van Hulst, W. P. M. H. Heemels, D. J. Antunes
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2504.05978v4 Announce Type: replace Abstract: Reinforcement learning (RL) is a powerful framework for decision-making in uncertain environments, but it often requires large amounts of data to learn an optimal policy. We address this challenge by incorporating prior model knowledge to guide exp...
118. An Automated Pipeline for Few-Shot Bird Call Classification: A Case Study with the Tooth-Billed Pigeon ​
Author: Abhishek Jana, Moeumu Uili, James Atherton, Mark O'Brien, Joe Wood, Leandra Brickson
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.SD
arXiv:2504.16276v3 Announce Type: replace Abstract: This paper presents a largely automated one-shot bird call classification pipeline, incorporating targeted manual quality control steps, designed for rare species absent from large publicly available classifiers like BirdNET and Perch. While these ...
119. SPD Matrix Learning for Neuroimaging Analysis: Perspectives, Methods, and Challenges ​
Author: Ce Ju, Reinmar Kobler, Antoine Collas, Motoaki Kawanabe, Cuntai Guan, Bertrand Thirion
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.IV, q-bio.NC
arXiv:2504.18882v3 Announce Type: replace Abstract: Neuroimaging provides essential tools for characterizing brain activity, structure, and connectivity through modalities that capture complementary aspects of brain organization. Across these diverse modalities, a unifying perspective arises when me...
120. Reinforcing Multi-Turn Reasoning in LLM Agents via Fine-Grained Reward Structure and Credit Assignment ​
Author: Quan Wei, Siliang Zeng, Chenliang Li, Zhongruo Wang, William Brown, Oana Frunza, Wei Deng, Anderson Schneider, Yuriy Nevmyvaka, Yang Katie Zhao, Alfredo Garcia, Mingyi Hong
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2505.11821v3 Announce Type: replace Abstract: Reinforcement Learning (RL) approaches have been wildly used to enhance the reasoning capabilities of Large Language Model (LLM) agents in long-horizon, multi-turn scenarios. Such interactions can be formalized as turn-level Markov decision process...
121. AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning ​
Author: Can Jin, Yang Zhou, Qixin Zhang, Hongwu Peng, Di Zhang, Zihan Dong, Marco Pavone, Ligong Han, Zhang-Wei Hong, Tong Che, Dimitris N. Metaxas
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2508.14313v4 Announce Type: replace Abstract: Test-time scaling strategies for Large Language Models predominantly rely on either reinforcement learning with sparse outcome rewards or search-based methods guided by static Process Reward Models. However, outcome-based RL often suffers from trai...
122. HIP: Hessian Interatomic Potentials without derivatives ​
Author: Andreas Burger, Luca Thiede, Nikolaj R{\o}nne, Varinia Bernales, Nandita Vijaykumar, Tejs Vegge, Arghya Bhowmik, Alan Aspuru-Guzik
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph, physics.comp-ph
arXiv:2509.21624v4 Announce Type: replace Abstract: Molecular Hessians, the second derivatives of the potential energy, are fundamental to many workflows in computational chemistry. Usually, accurate Hessians are computationally expensive to calculate and scale poorly with system size, whether compu...
123. Perseus: Interactive Time Series Segmentation with Sparse Supervision via Stateful Memory ​
Author: Ching Chang, Ming-Chih Lo, Chiao-Tung Chan, Wen-Chih Peng, Tien-Fu Chen
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.09930v2 Announce Type: replace Abstract: Real-world systems, ranging from industrial manufacturing to wearable healthcare, generate multivariate time series with hierarchical states ranging from coarse regimes to fine-grained events. Unlike zero- or few-shot segmentation, our setting uses...
124. Doctor Rashomon and the UNIVERSE of Madness: Variable Importance with Unobserved Confounding and the Rashomon Effect ​
Author: Jon Donnelly, Srikar Katta, Emanuele Borgonovo, Cynthia Rudin
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.12734v2 Announce Type: replace Abstract: Variable importance (VI) methods are often used for hypothesis generation, feature selection, and scientific validation. In the standard VI pipeline, an analyst estimates VI for a single predictive model with only the observed features. However, th...
125. LTR-ICD: A Ranking-Aware Framework for Automatic ICD Coding ​
Author: Mohammad Mansoori, Amira Soliman, Farzaneh Etminani
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IR
arXiv:2510.13922v2 Announce Type: replace Abstract: Clinical notes contain unstructured text provided by clinicians during patient encounters. These notes are usually accompanied by a sequence of diagnostic codes following the International Classification of Diseases (ICD). Correctly assigning and o...
126. Benchmarking noisy label detection methods ​
Author: Henrique Pickler, Jorge K. S. Kamassury, Danilo Silva
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2510.16211v2 Announce Type: replace Abstract: Label noise is a common problem in real-world datasets, affecting both model training and validation. Clean data are essential for achieving strong performance and ensuring reliable evaluation. While various techniques have been proposed to detect ...
127. BIPPO: Budget-Aware Independent PPO for Energy-Efficient Federated Learning Services ​
Author: Anna Lackinger, Andrea Morichetta, Pantelis A. Frangoudis, Schahram Dustdar
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.MA
arXiv:2511.08142v2 Announce Type: replace Abstract: Federated Learning (FL) is a promising machine learning solution in large-scale IoT systems, guaranteeing load distribution and privacy. However, FL does not natively consider infrastructure efficiency, a critical concern for systems operating in r...
128. Spatially Aware Dictionary-Free Koopman Eigenfunction Identification for Modeling and Control ​
Author: David Grasev
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.22648v2 Announce Type: replace Abstract: A spatially aware dictionary-free eigenfunction discovery (SADFED) framework is proposed for identification of low-rank Koopman models from data without prescribing a lifting dictionary, kernel, or neural-network eigenfunction architecture. A refer...
129. Efficient Inference for Inverse Reinforcement Learning and Dynamic Discrete Choice Models ​
Author: Lars van der Laan, Aur'elien Bibaut, Nathan Kallus
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.TH
arXiv:2512.24407v2 Announce Type: replace Abstract: In many sequential decision-making problems, researchers observe actions but not the rewards that drive behavior, yet still wish to evaluate and compare counterfactual policies. Inverse reinforcement learning (IRL) and dynamic discrete choice (DDC)...
130. When to Ponder: Adaptive Compute Allocation for Code Generation via Test-Time Training ​
Author: Gihyeon Sim
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2601.00894v2 Announce Type: replace Abstract: Large language models apply uniform computation to all inputs, regardless of difficulty. We propose PonderTTT, a gating strategy using the TTT layer's self-supervised reconstruction loss to selectively trigger Test-Time Training (TTT) updates. The ...
131. AgentOCR: Reimagining Agent History via Optical Self-Compression ​
Author: Lang Feng, Fuchao Yang, Feng Chen, Xin Cheng, Haiyang Xu, Zhenglin Wan, Ming Yan, Bo An
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2601.04786v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) enable agentic systems trained with reinforcement learning (RL) over multi-turn interaction, but practical deployment is bottlenecked by rapidly growing textual histories that inflate token and memory...
132. GroupSegment-SHAP: Shapley Value Explanations with Group-Segment Players for Multivariate Time Series ​
Author: Jinwoong Kim, Sangjin Park
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GT
arXiv:2601.06114v2 Announce Type: replace Abstract: Multivariate time-series models achieve strong predictive performance in healthcare, industry, energy, and finance, but how they combine cross-variable interactions with temporal dynamics remains unclear. SHapley Additive exPlanations (SHAP) are wi...
133. Generalization Measures under Controlled Covariate Shift: A Regime-Aware Benchmark ​
Author: Sora Nakai, Youssef Fadhloun, Kacem Mathlouthi, Kotaro Yoshida, Ganesh Talluri, Ioannis Mitliagkas, Hiroki Naganuma
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.01718v2 Announce Type: replace Abstract: Predicting generalization from quantities available before target-test evaluation remains a central challenge in deep learning. The systematic benchmark of Jiang et al. (2020) evaluated many generalization measures, but it focused on independent an...
134. Interpretability in Deep Time Series Models Demands Semantic Alignment ​
Author: Giovanni De Felice, Riccardo D'Elia, Alberto Termine, Pietro Barbiero, Giuseppe Marra, Silvia Santini
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.02239v3 Announce Type: replace Abstract: Deep time series models continue to improve predictive performance, yet their deployment remains limited by their black-box nature. In response, existing interpretability approaches in the field keep focusing on explaining the internal model comput...
135. Maximum Likelihood Reinforcement Learning ​
Author: Fahim Tajwar, Guanning Zeng, Yueer Zhou, Yuda Song, Daman Arora, Yiding Jiang, Jeff Schneider, Ruslan Salakhutdinov, Haiwen Feng, Andrea Zanette
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.02710v3 Announce Type: replace Abstract: Reinforcement learning (RL) is the method of choice for training models in setups where the objective function can only be evaluated by sampling from the model. Our key observation is that when the feedback is terminal and binary, models implicitly...
136. Degree-Mass Message Passing for Betweenness Ranking in Directed and Undirected Networks ​
Author: Justin Dachille, Aurora Rossi, Sunil Kumar Maurya, Frederik Mallmann-Trenn, Xin Liu, Fr'ed'eric Giroire, Tsuyoshi Murata, Emanuele Natale
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.09716v2 Announce Type: replace Abstract: Computing the importance of nodes in networks is a long-standing fundamental problem that has driven extensive study of various centrality measures. A particularly well-known centrality measure is betweenness centrality, whose exact computation bec...
137. Interpretable clustering via optimal multi-way decision trees ​
Author: Hayato Suzuki, Shunnosuke Ikeda, Naoki Nishimura, Yuichi Takano
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.13586v2 Announce Type: replace Abstract: Clustering is a fundamental unsupervised learning technique for uncovering data structures to facilitate knowledge discovery and decision-making. While clustering accuracy is crucial, interpretability significantly impacts the practical value of cl...
138. Investigating Target Class Influence on Neural Network Compressibility for Energy-Autonomous Avian Monitoring ​
Author: Nina Brolich, Simon Geis, Maximilian Kasper, Alexander Barnhill, Axel Plinge, Dominik Seu{\ss}
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.17751v2 Announce Type: replace Abstract: Biodiversity loss poses a significant threat to humanity, making wildlife monitoring essential for assessing ecosystem health. Avian species are ideal subjects for this due to their popularity and the ease of identifying them through their distinct...
139. Efficient Exploration at Scale ​
Author: Seyed Mohammad Asghari, Chris Chute, Vikranth Dwaracherla, Xiuyuan Lu, Mehdi Jafarnia, Victor Minden, Zheng Wen, Benjamin Van Roy
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.17378v2 Announce Type: replace Abstract: We develop an online learning algorithm that dramatically improves the data efficiency of reinforcement learning from human feedback (RLHF). Our algorithm incrementally updates reward and language models as choice data is received. The reward model...
140. Calibrate-Then-Delegate: Safety Monitoring with Risk and Budget Guarantees via Model Cascades ​
Author: Edoardo Pona, Milad Kazemi, Mehran Hosseini, Yali Du, David Watson, Osvaldo Simeone, Nicola Paoletti
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.14251v2 Announce Type: replace Abstract: Monitoring LLM safety at scale requires balancing cost and accuracy: a cheap latent-space probe can screen every input, but hard cases should be escalated to a more expensive expert. Existing cascades delegate based on probe uncertainty, but uncert...
141. AutoOR: Scalably Post-training LLMs to Autoformulate Operations Research Problems ​
Author: Sumeet Ramesh Motwani, Chuan Du, Aleksander Petrov, Christopher Davis, Philip Torr, Antonio Papania-Davis, Weishi Yan
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2604.16804v4 Announce Type: replace Abstract: Optimization problems are central to decision-making in manufacturing, logistics, scheduling, and other industrial settings. Translating complicated descriptions of these problems into solver-ready formulations requires specialized operations resea...
142. RefusalGuard: Geometry-Preserving Fine-Tuning for Safety in LLMs ​
Author: Sadia Asif, Mohammad Mohammadi Amiri
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE, cs.CL, cs.CR
arXiv:2605.01913v2 Announce Type: replace Abstract: Fine-tuning safety-aligned language models for downstream tasks often leads to substantial degradation of refusal behavior, making models vulnerable to adversarial misuse. While prior work has shown that safety-relevant features are encoded in stru...
143. Preference-Based Self-Distillation: Beyond KL Matching via Reward Regularization ​
Author: Xin Yu, Liuchen Liao, Yiwen Zhang, Yingchen Yu, Lingzhou Xue, Qinzhen Guo
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.05040v2 Announce Type: replace Abstract: On-policy distillation is an efficient alternative to reinforcement learning, offering dense token-level training signals. However, its reliance on a stronger external teacher has driven recent work on on-policy self-distillation, where the same mo...
144. GRALIS: Fusing Coalition and Gradient Attribution with Closed-Form Conservation Error and Finite-Sample Guarantees ​
Author: Raimondo Fanale
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2605.05480v3 Announce Type: replace Abstract: The main post-hoc XAI methods for deep networks -- GradCAM, SHAP, LIME, Integrated Gradients -- originate from heterogeneous theoretical foundations and are not naturally comparable within a single representation. A recent benchmark also finds thei...
145. Learning Minimal-Deviation Corrections for Multi-Dimensional Mismodelling in HEP Simulations ​
Author: Matthias Schott, Lucie Flek
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, hep-ex
arXiv:2605.07460v2 Announce Type: replace Abstract: Accurate Monte Carlo (MC) modelling in high-energy physics is challenging, particularly in complex scenarios where simulations fail to reproduce observed data. In practice, experimental information is often limited to one-dimensional (1D) distribut...
146. STS: Efficient Sparse Attention with Speculative Token Sparsity ​
Author: Jiangnan Yu, Ceyu Xu, Yongji Wu, Yuan Xie
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2605.15508v3 Announce Type: replace Abstract: The quadratic complexity of attention imposes severe memory and computational bottlenecks on Large Language Model (LLM) inference. This challenge is particularly acute for emerging agentic applications that require processing multi-million token se...
147. Behavior-Consistent Deep Reinforcement Learning ​
Author: Marcel Hussing, Liv G. d'Aliberti, Claas Voelcker, Benjamin Eysenbach, Eric Eaton
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.21214v3 Announce Type: replace Abstract: Reinforcement learning (RL) often exhibits high variance across training runs, leading to unreliable performance and posing a major challenge to deployment in real-world domains. In this work, we address the challenge of cross-run policy divergence...
148. The Fast Mixing Mechanism for Differential Privacy ​
Author: Omri Lev, Moshe Shenfeld, Vishwak Srinivasan, Katrina Ligett, Ashia C. Wilson
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2605.30600v2 Announce Type: replace Abstract: Randomized sketching is a central tool for compressing large-scale optimization problems while preserving accuracy. In particular, sketches that are based on structured matrices, such as the Hadamard matrix, can be applied efficiently and often yie...
149. PROBE-Web: An Interactive System for Probing Evaluation Landscapes of Knowledge Graph Completion Models ​
Author: Sooho Moon, Yunyong Ko
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.08926v3 Announce Type: replace Abstract: Knowledge graph completion (KGC) models are commonly evaluated using rank-based metrics such as MRR and Hits@K, despite different users often requiring different evaluation perspectives. In this demo, we present PROBE-Web, an interactive system for...
150. INFUSER: Influence-Guided Self-Evolution Improves Reasoning ​
Author: Siyu Chen, Miao Lu, Beining Wu, Heejune Sheen, Fengzhuo Zhang, Shuangning Li, Zhiyuan Li, Jose Blanchet, Tianhao Wang, Zhuoran Yang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.GT, stat.ML
arXiv:2606.09052v4 Announce Type: replace Abstract: Self-evolution offers a scalable path to stronger reasoning: a pretrained language model improves itself with only minimal external supervision. Yet existing methods either depend on extensively curated or teacher-generated training data, or, when ...
151. Uncertainty-aware Multi-fidelity Closure via Conditional Normalizing Flows ​
Author: Jice Zeng, Shady E. Ahmed, David Barajas-Solano, Panos Stinis
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2606.09857v2 Announce Type: replace Abstract: Reduced-order models (ROMs) provide efficient surrogates for complex multiscale systems, but their predictive accuracy is often compromised by truncation errors and the inadequate representation of interactions between resolved and unresolved scale...
152. Detecting Functional Memorization in Code Language Models ​
Author: Matthieu Meeus, Anil Ramakrishna, Shengyuan Hu, Matthew Grange, Zheng Xu, Luca Melis
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CR
arXiv:2606.12764v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to generate code at scale. Meanwhile, prior work has investigated whether training data may be recoverable from model outputs, by auditing the textual overlap between training examples and model ge...
153. What a World Model Represents Is Three Questions ​
Author: Donna Vakalis
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.06640v2 Announce Type: replace Abstract: World models learn task-relevant information through many routes: observation reconstruction, recurrent state, temporal filtering, and explicit task supervision. Different routes can make different variables available. The same variable can also be...
154. An Introduction to Bayesian and Frequentist Simulation-Based Inference with Machine Learning ​
Author: Maximilian Dax, Theo Heimel, Gilles Louppe
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.CO, astro-ph.GA, hep-ex, hep-ph, stat.ML
arXiv:2607.21702v2 Announce Type: replace Abstract: Simulation-based inference (SBI) with machine learning is an increasingly important tool for solving inverse problems in science and engineering, including parameter inference and the inversion of detector effects. We provide an overview of the Bay...
155. SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation ​
Author: Wen Wang, Jiahua Bao, Tu Yongsiqi, Yihao Liu, Haotian Zhou, Haoxuan Ma, Mengyu Zhou, Wenkui Fan, Junwei He, Xiaoxi Jiang, Guanjun Jiang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.03092v2 Announce Type: replace Abstract: We aim to improve model performance in multi-reward reinforcement learning training process. Existing Group reward-Decoupled Normalization Policy Optimization (GDPO) has mitigated the issue of reward signals masking one another during direct scalar...
156. ELVAE: Evidential Learning-Based Variational Autoencoder for Uncertainty-Aware Generation ​
Author: Ge Wang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.10398v2 Announce Type: replace Abstract: ELVAE places an input-dependent normal--inverse-gamma (NIG) hierarchy at each VAE latent coordinate, separating location uncertainty $u_{\mathrm{epi}}=\beta/[\nu(\alpha-1)]$ from conditional variability $u_{\mathrm{var}}=\beta/(\alpha-1)$. The marg...
157. When Does Forecasting Reveal Temporal Structure? A Stability Analysis of Time-Series Structural Selection ​
Author: Qipeng Qian, Yuntao Qian
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.10433v5 Announce Type: replace Abstract: Forecast accuracy is often used as a proxy for temporal structure discovery, but predictive performance and structural identifiability are not equivalent. Different temporal mechanisms can achieve similar forecast errors, while small forecast diffe...
158. Towards Truly Unsupervised Evaluation of Feature Selection ​
Author: Hafiz Saud Arshad, Muhammad Rajabinasab, Arthur Zimek
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12057v2 Announce Type: replace Abstract: Feature selection is one of the most important and fundamental tasks in data mining, tackled by a family of methods with an established set of evaluation techniques to measure the quality of a specific method. Most of the methods commonly used for ...
159. Information Geometry of Message Passing ​
Author: Mykola Lukashchuk, Kyrylo Yemets, Alex Ledbetter, .{I}smail \c{S}en"oz
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15922v2 Announce Type: replace Abstract: We show that the natural-gradient stationary condition of variational inference has an edge-local form on a Forney-style factor graph. We start from the Bethe free energy and constrain a selected edge marginal to an exponential family. At a station...
160. GEO-Flag: Detecting and Measuring GEO-Optimized Web Content ​
Author: Junjie Chu, Ye Leng, Mingjie Li, Yun Shen, Xinyue Shen, Yang Zhang
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.IR
arXiv:2608.16824v2 Announce Type: replace Abstract: Generative Engine Optimization (GEO) modifies web content to increase its likelihood of being selected and cited by generative search engines. This can give strategically optimized pages visibility disproportionate to their authority or relevance a...
161. Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements ​
Author: Zhi Zheng, Rongsheng Chen, Yunpeng Ba, Zhenkun Wang, Yee Whye Teh, Wee Sun Lee
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17310v2 Announce Type: replace Abstract: Reinforcement Learning (RL) has been promising in single-turn LLM fine-tuning. However, long-horizon agentic reasoning introduces increasingly branching interactions and sparse rewards, exposing several limitations of RL: its heavyweight backpropag...
162. Separating Covariate Shift from Mechanism Change with Two Discriminators: CJSD, a Conditional Discrepancy with an Exact Covariate-Concept Decomposition ​
Author: Kentaro Oda
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.19885v2 Announce Type: replace Abstract: After the inputs X are known, how much additional information does the label Y carry about which dataset a sample came from? That single quantity -- estimable as the difference of two discriminators' held-out cross-entropies, D_CJS = CE(Z|X) - CE(Z...
163. Green BOA: Determining the environmental break-even point for ML-based data compression ​
Author: Caterina Doglioni, Thomas Elliott, Akshat Gupta, Hanzila Hussain, Sanjiban Sengupta, Zhengkai Sun
Published: 8/24/2026, 4:00:00 AM
Categories: cs.LG, hep-ex, physics.comp-ph
arXiv:2608.19994v2 Announce Type: replace Abstract: We summarise the outcome of two summer internship projects based at the University of Manchester, focused on the break-even point in terms of environmental sustainability for ML-based data compression algorithms. Using the example of a ML-based los...
164. On the Within-class Variation Issue in Alzheimer's Disease Detection ​
Author: Jiawen Kang, Dongrui Han, Lingwei Meng, Jingyan Zhou, Jinchao Li, Xixin Wu, Helen Meng
Published: 8/24/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.CL, cs.LG, cs.SD, q-bio.NC
arXiv:2409.16322v4 Announce Type: replace-cross Abstract: Alzheimer's Disease (AD) detection commonly employs machine learning classification models to distinguish between individuals with AD and those without. Different from conventional classification tasks, AD detection involves substantial withi...
165. DAOP: Data-Aware Offloading and Predictive Pre-Calculation for Efficient MoE Inference ​
Author: Yujie Zhang, Shivam Aggarwal, Tulika Mitra
Published: 8/24/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2501.10375v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) models, though highly effective for various machine learning tasks, face significant deployment challenges on memory-constrained devices. While GPUs offer fast inference, their limited memory compared to CPUs means no...
166. The Intrinsic Dimension of Prompts in Internal Representations of Large Language Models ​
Author: Karthik Viswanathan, Yuri Gardinazzi, Giada Panerai, Alberto Cazzaniga, Matteo Biagetti
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2501.10573v2 Announce Type: replace-cross Abstract: We study the geometry of token representations at the prompt level in large language models through the lens of intrinsic dimension. Viewing transformers as mean-field particle systems, we estimate the intrinsic dimension of the empirical mea...
167. Regression-Based Estimation of Causal Effects in the Presence of Selection Bias and Confounding ​
Author: Marlies Hafer, Alexander Marx
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2503.20546v2 Announce Type: replace-cross Abstract: We consider the problem of estimating the expected causal effect $E[Y|do(X)]$ for a target variable $Y$ when treatment $X$ is set by intervention, focusing on continuous random variables. In settings without selection bias or confounding, $E[...
168. Explaining Intrinsic Moral Self-Correction with Mechanistic Interpretability ​
Author: Yu-Ting Lee, Fu-Chieh Chang, Yu-En Shu, Hui-Ying Shih, Pei-Yuan Wu
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2505.11924v4 Announce Type: replace-cross Abstract: Intrinsic moral self-correction refers to the phenomenon where a language model refines its ethical judgments or aligns its outputs purely through prompting. While effective across diverse tasks, its mechanism remains unclear. We hypothesize ...
169. CPC-CMS: Cognitive Pairwise Comparison Classification Model Selection Framework for Document-level Sentiment Analysis ​
Author: Jianfei Li, Kevin Kam Fung Yuen
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2507.14022v3 Announce Type: replace-cross Abstract: This study proposes the Cognitive Pairwise Comparison Classification Model Selection (CPC-CMS) framework for document-level sentiment analysis. The CPC, based on expert knowledge judgment, is used to calculate the weights of evaluation criter...
170. Query Efficient Structured Matrix Learning ​
Author: Noah Amsel, Pratyush Avi, Tyler Chen, Feyza Duman Keles, Chinmay Hegde, Cameron Musco, Christopher Musco, David Persson
Published: 8/24/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, cs.NA, math.NA
arXiv:2507.19290v2 Announce Type: replace-cross Abstract: We study the problem of learning a structured approximation (low-rank, sparse, banded, etc.) to an unknown matrix $A$ given access to matrix-vector product (matvec) queries of the form $x \rightarrow Ax$ and $x \rightarrow A^Tx$. This problem...
171. MeltwaterBench: Deep learning for spatiotemporal downscaling of surface meltwater ​
Author: Bj"orn L"utjens, Patrick Alexander, Raf Antwerpen, Til Widmann, Guido Cervone, Marco Tedesco
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, physics.ao-ph, physics.data-an
arXiv:2512.12142v2 Announce Type: replace-cross Abstract: The Greenland ice sheet is melting at an accelerated rate due to processes that are not fully understood and hard to measure. The distribution of surface meltwater can help understand these processes and is observable through remote sensing, ...
172. Actively Learning Joint Contours of Multiple Computer Experiments ​
Author: Shih-Ni Prim, Kevin R. Quinlan, Paul Hawkins, Jagadeesh Movva, Annie S. Booth
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ME, cs.LG
arXiv:2512.13530v2 Announce Type: replace-cross Abstract: Contour location---the process of sequentially training a surrogate model to identify the design inputs that result in a pre-specified response value from a single computer experiment---is a well-studied active learning problem. Here, we tack...
173. Deterministic and probabilistic neural surrogates of global hybrid-Vlasov simulations ​
Author: Daniel Holmberg, Ivan Zaitsev, Markku Alho, Ioanna Bouri, Fanni Franssila, Haewon Jeong, Minna Palmroth, Teemu Roos
Published: 8/24/2026, 4:00:00 AM
Categories: physics.space-ph, cs.LG, physics.plasm-ph
arXiv:2601.12614v4 Announce Type: replace-cross Abstract: Hybrid-Vlasov simulations resolve ion-kinetic effects in the solar wind-magnetosphere interaction, but even 5D (2D + 3V) configurations are computationally expensive. We show that graph-based machine learning emulators can learn the spatiotem...
174. CFM: Language-aligned Concept Foundation Model for Vision ​
Author: Kai Wittenmayer, Sukrut Rao, Amin Parchami-Araghi, Bernt Schiele, Jonas Fischer
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2601.13798v3 Announce Type: replace-cross Abstract: Language-aligned vision foundation models perform strongly across diverse downstream tasks. Yet, their learned representations remain opaque, making interpreting their decision-making difficult. Recent work decompose these representations int...
175. SEISMO: Explanation-Aware, Trajectory-Conditioned LLM Agents for Sample-Efficient Molecular Optimisation ​
Author: Fabian P. Kr"uger, Andrea Hunklinger, Adrian Wolny, Tim J. Adler, Igor Tetko, Santiago David Villalba
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-bio.BM
arXiv:2602.00663v3 Announce Type: replace-cross Abstract: Optimizing molecules to achieve desired properties is a central bottleneck across the chemical sciences, particularly in the pharmaceutical industry, where it underlies the discovery of new drugs. Since molecular property evaluation often rel...
176. Infinite-dimensional generative diffusions via Doob's h-transform ​
Author: Thorben Pieper-Sethmacher, Daniel Paulin
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2602.06621v2 Announce Type: replace-cross Abstract: This paper introduces a rigorous framework for defining generative diffusion models in infinite dimensions via Doob's h-transform. Rather than relying on time reversal of a noising process, a reference diffusion is forced towards the target d...
177. A Deep Reinforcement Learning Framework for Closed-loop Guidance of Fish Schools via Virtual Agents ​
Author: Takato Shibayama, Hiroaki Kawashima
Published: 8/24/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, q-bio.PE
arXiv:2603.28200v2 Announce Type: replace-cross Abstract: Guiding collective motion in biological groups is a fundamental challenge in understanding social interaction rules. In this study, we propose a deep reinforcement learning (RL) framework for closed-loop guidance of fish schools using virtual...
178. Optimistic Online LQR via Intrinsic Rewards ​
Author: Marcell Bartos, Bruce D. Lee, Lenart Treven, Andreas Krause, Florian D"orfler, Melanie N. Zeilinger
Published: 8/24/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC
arXiv:2603.28938v2 Announce Type: replace-cross Abstract: Optimism in the face of uncertainty is a popular approach to balance exploration and exploitation in reinforcement learning. Here, we consider the online linear quadratic regulator (LQR) problem, i.e., to learn the LQR corresponding to an unk...
179. Automatic classification pipeline for glitches in the Virgo detector ​
Author: Tiago Fernandes, Francesco Di Renzo, Antonio Onofre, Alejandro Torres-Forn'e, Jos'e A. Font
Published: 8/24/2026, 4:00:00 AM
Categories: gr-qc, astro-ph.IM, cs.LG
arXiv:2604.13687v2 Announce Type: replace-cross Abstract: Glitches frequently contaminate data in gravitational-wave detectors, complicating the observation and analysis of astrophysical signals. This work introduces VIGILant, an automatic pipeline for classification and visualization of glitches in...
180. Compared to What? Baselines and Metrics for Counterfactual Prompting ​
Author: Zihao Yang, Mosh Levy, Yoav Goldberg, Byron C. Wallace
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2605.01048v2 Announce Type: replace-cross Abstract: Counterfactual prompting (i.e., perturbing a single factor and measuring output change) is widely used to evaluate things like LLM bias and CoT faithfulness. But in this work we argue that observed effects cannot be attributed to the targeted...
181. RouteScan: A Non-Intrusive Approach to Auditing MoE LLMs Safety via Expert Routing Telemetry ​
Author: Bo Lv, Zhiheng Xu, KeDong Xiu, Ruyi Ding, Tianhang Zheng, Zhibo Wang, Kui Ren
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CR, cs.AR, cs.CL, cs.LG
arXiv:2605.24817v2 Announce Type: replace-cross Abstract: As Mixture-of-Experts (MoE) architectures are increasingly adopted for scaling Large Language Models (LLMs), safety auditing becomes necessary to verify whether these models produce or facilitate harmful behaviors during operation. However, e...
182. On Finite-sample Concentration of Median of Incomplete U-Statistics ​
Author: Nong Minh Hieu, Antoine Ledent
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2606.00661v2 Announce Type: replace-cross Abstract: Median-of-means (MoM) is a powerful technique that theoretically enables near sub-Gaussian finite-sample rate for parameter estimation when the underlying data distribution is heavy-tailed (e.g., assumed to have only two first finite moments)...
183. Self-Revising Discovery Systems for Science: A Categorical Framework for Agentic Artificial Intelligence ​
Author: Fiona Y. Wang, Markus J. Buehler
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cond-mat.mtrl-sci, cs.CL, cs.LG, math.CT
arXiv:2606.01444v2 Announce Type: replace-cross Abstract: Scientific discovery is not only answer generation but revision of the representational regime in which evidence, artifacts, operations, and verifiers are typed. We develop a category-theoretic account of agentic discovery for materials scien...
184. Towards Automated Discovery: A Review of Generative Models, Multimodal Learning and Closed-Loop Workflows in Inverse Materials Design ​
Author: Anand Babu, Rog'erio Almeida Gouv^ea, Gian-Marco Rignanese
Published: 8/24/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.ET, cs.LG, physics.app-ph, physics.comp-ph
arXiv:2606.02507v2 Announce Type: replace-cross Abstract: Inverse materials design is shifting materials discovery from forward prediction toward targeted proposal of candidates that satisfy objectives under physical constraints. Here, we review advances in generative crystal structure modeling, mul...
185. Adaptive Inference for Resource-Constrained Dynamic Pricing ​
Author: Ruicheng Ao, Jiashuo Jiang, David Simchi-Levi
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2606.03736v2 Announce Type: replace-cross Abstract: We study resource-constrained dynamic pricing when the seller seeks revenue and valid inference about demand at a price fixed before the selling season. Depletion can remove every feasible price near the target, so randomization over the rema...
186. Valid Inference with Synthetic Data via Task Exchangeability ​
Author: Lezhi Tan, Tijana Zrnic
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ME, cs.AI, cs.LG, stat.ML
arXiv:2606.13629v2 Announce Type: replace-cross Abstract: There is a proliferation of work arguing for the use of synthetic data in scientific research. For example, social scientists are arguing for the use of LLM-generated "silicon samples" in pilot studies; AI evaluations increasingly rely on "LL...
187. The Metanym Game: An LLM Benchmark Without Ground Truth That Rises With the Models It Measures ​
Author: David Nordfors
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2606.21008v3 Announce Type: replace-cross Abstract: We present evidence that analogy is at the core of LLM intelligence. In our benchmark, LLMs compete in generating sets of analogous statements and rate each other's sets on their own understandings of factual correctness, beauty, intelligence...
188. FeLoG: Scalable and Efficient Distributed Graph Embedding with Feedback Loop Mechanism ​
Author: Peng Fang, Arijit Khan, Ziqiang Wu, Zhenli Li, Yibo Zhou, Fang Wang, Dan Feng
Published: 8/24/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2606.22180v3 Announce Type: replace-cross Abstract: Graph embedding maps graph nodes into low-dimensional vectors to support applications such as recommendation, fraud detection, and graph-based retrieval-augmented generation (GraphRAG). As graphs scale to billions of edges, scalable and effic...
189. The Dual Nature of LLM Persona: Aggregated Tendencies and Frame-Dependent Geometry ​
Author: Yuan Yuan
Published: 8/24/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, math.DG
arXiv:2607.02368v2 Announce Type: replace-cross Abstract: Evaluations of LLM personas via psychometric questionnaires typically rely on aggregate scores, discarding within-instance correlation structure. We test whether this geometric structure is intrinsic or frame-dependent. Constructing within-in...
190. Latent Softmax for Data-Efficient Phoneme-Based Multilingual ASR Across Tonal and Non-Tonal Languages ​
Author: Saierdaer Yusuyin, Nanling Jiang, Hao Huang, Zhijian Ou
Published: 8/24/2026, 4:00:00 AM
Categories: eess.AS, cs.LG
arXiv:2608.01281v2 Announce Type: replace-cross Abstract: Phoneme-based multilingual automatic speech recognition (ASR) can share acoustic evidence across languages more directly than language-specific subword modeling. When tonal and non-tonal languages are jointly trained, however, their supervisi...
191. A continually expandable foundation model for brain MRI ​
Author: Michail Mamalakis, Carmen Jimenez-Mesa, Yonghao Li, Hao Chen, Chao Li, Antonios Mamalakis, John Suckling, Richard Bethlehem, Stephen J. Price, Richard J. Gilbertson, Pietro Lio
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.08319v3 Announce Type: replace-cross Abstract: Brain magnetic resonance imaging (MRI) is central to neuroscience and clinical assessment, but models are commonly developed for individual diseases, populations or imaging protocols. Foundation models promise more general representations, ye...
192. Defining Decentralization: An Ontological Perspective ​
Author: Jakub Kacper Szel\k{a}g, Aydin Abadi, Mohammad Naseri
Published: 8/24/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG, cs.LO, cs.SY, eess.SY
arXiv:2608.09748v2 Announce Type: replace-cross Abstract: Decentralization as a concept in computer science has existed for over half a century. Despite its fundamental role across domains such as security, distributed computing, artificial intelligence, cloud infrastructures, and Internet of Things...
193. Attributing Preprocessing Invariance in Spectral Foundation Models ​
Author: Dongjun Wei, Hongyi Wu, Yinuo Zou
Published: 8/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CE, cs.LG
arXiv:2608.14227v2 Announce Type: replace-cross Abstract: Preprocessing invariance is an appealing goal for spectral foundation models: a frozen model should remain useful when laboratories preprocess spectra differently. It is usually measured by training a classifier under one preprocessing pipeli...
194. Non-Shattering at and Above the Dynamical Temperature in the Spherical Pure p-Spin Model ​
Author: Taegyun Kim
Published: 8/24/2026, 4:00:00 AM
Categories: math.PR, cond-mat.dis-nn, cs.LG, math-ph, math.MP
arXiv:2608.14369v2 Announce Type: replace-cross Abstract: We consider the notion of shattering introduced by Ben Arous and Jagannath for spherical pure $p$-spin glasses with overlap $q$. For every $p\geq 3$ and $0\sqrt{(p-2)/(p-1)}$. The proof combines a deterministic $N+1$ bound for disjoint bands ...
195. Mint-Agent: Introducing Finance-Native Agentic Foundation Models ​
Author: Agent Team, Kun Wang, Gavin Zhang, Yaze Geng, Lei Tang, Yaoyang Yi, Zonghan Wu, Yifan Hu, Qingsong Wen, Yilei Shao
Published: 8/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.16386v2 Announce Type: replace-cross Abstract: Financial agents must do more than recall domain knowledge: they must be both reliable, executing precise operations over grounded evidence, and executive, sustaining long-horizon research whose conclusions remain auditable. We present Mint-A...