Skip to content

arXiv cs.LG - 2026-08-25 ​

485 items collected.


1. Model of Models: When Does Emitting a Specialist Beat Attending, Adapting, or Tuning? ​

Author: John C. Howell
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21386v1 Announce Type: new Abstract: Given a task described by a few examples, how should a model be specialized to it? Four mechanisms are available -- zero-shot, in-context attention, test-time gradient adaptation, and emitting specialist weights from a hypernetwork -- yet the operating...

📖 Read original article


2. Runtime Action Interference for AI Control of AlphaStar in StarCraft II ​

Author: Jaymari Chua, Chen Wang, Liming Zhu, Lina Yao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY, cs.HC, cs.MA

arXiv:2608.21398v1 Announce Type: new Abstract: A trained reinforcement learning policy does not determine the complete behavior that users encounter: deployment code still schedules, admits, suppresses, or replaces its proposed actions. We contribute \emph{runtime action interference} (RAI), an AI ...

📖 Read original article


3. Federated Ensemble Forecasting Under Supply-Chain Market Volatility ​

Author: Shunmukha Sagar Puppala
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21399v1 Announce Type: new Abstract: Supply chain forecasting systems increasingly operate under market shocks, non-identically distributed regional demand, and limited willingness to centralize commercial data. This work proposes Federated Ensemble Forecasting with Negative-Correlation L...

📖 Read original article


4. Class-Conditioned Gaussian Mixture Modeling for Imbalanced Time Series Quantification ​

Author: Md Shahriar Kabir, Mayesha Maliha R. Mithila, Anne H. H. Ngu, Myl`ene C. Q. Farias, Byron Gao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21473v1 Announce Type: new Abstract: Quantification, estimating class prevalences in bags of unlabeled instances is vital in domains where aggregate statistics are more important than individual instance labels, such as biosignal monitoring, fall detection, and activity recognition. We in...

📖 Read original article


5. Congruence Decomposition with Neural Block Solvers for Large-Scale PCI Assignment ​

Author: Yeqing Qiu, Chengpiao Huang, Ye Xue, Akang Wang, Fan Xu, Zhipeng Jiang, Dong Zhang, Ruoyu Sun, Qingjiang Shi, Zhi-Quan Luo
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2608.21485v1 Announce Type: new Abstract: Physical Cell Identity (PCI) assignment is essential for interference management in dense 5G networks. As cellular networks scale, PCI reuse becomes unavoidable, which may cause collisions, confusions, and multiple forms of modular interference. Jointl...

📖 Read original article


6. KAN-Robust-Bench: A Benchmark for Evaluating the Robustness of Kolmogorov-Arnold Networks ​

Author: Mohammad Meymani, Roozbeh Razavi-Far
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21488v1 Announce Type: new Abstract: While machine learning models have demonstrated strong performance in many domains, these models have shown profound vulnerabilities when they are exposed to adversarial threats. While adversarial attacks fall into various categories, the most prominen...

📖 Read original article


Author: Ricardo Fitas
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH

arXiv:2608.21496v1 Announce Type: new Abstract: AI systems increasingly generate alternatives, inspect evidence, and deploy a selected output. Validation is therefore target-relative: evidence certifies deployment only in directions resolved by the interventions that produced it. We represent valida...

📖 Read original article


8. Selection of Heart Sound Segments for Synchronous Classification of Multi-channel Heart Sounds ​

Author: Marcelo Nogueira, Jorge H. Oliveira, Carlos F. Ferreira, Miguel T. Coimbra, Al'ipio M. Jorge
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21499v1 Announce Type: new Abstract: Cardiac auscultation remains the most cost-effective screening procedure for cardiovascular diseases, and requires listening at the four main auscultation spots. Despite this, automatic heart sound analysis algorithms mostly classify patients using a s...

📖 Read original article


9. ChemDIRT: A Diversified Instruction, Representation, and Task Benchmark for Robust Chemistry-LLM Evaluation ​

Author: Eric Inae, Tim Gunn, Chris Bond, Meng Jiang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21504v1 Announce Type: new Abstract: The rapid advancement of large language models (LLMs) has led to increasing interest in their application to scientific domains such as chemistry. However, existing chemistry benchmarks often provide only a narrow view of model capability, focusing on ...

📖 Read original article


10. Multimodal Injury Risk and Performance Prediction in Tennis Using Weighted Ensemble Learning ​

Author: Weihao Qu, Dongyang Wang, Ling Zheng, Francisco E. Alvarez, Shobharani Polasa, Jiacun Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21530v1 Announce Type: new Abstract: Machine learning has had a positive impact on the sports industry, with one of its most promising applications being the prediction of athlete performance and injury risk. Recent advances have employed state-of-the-art models to improve prediction accu...

📖 Read original article


11. Federated Continual Learning as a Distributed Drift-Plus-Penalty Control Problem ​

Author: Nazreen Shah, Naveen Kumar Reddy Somireddy, Zubair Shaban, Ranjitha Prasad, B. N. Bharath
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21539v1 Announce Type: new Abstract: Federated Continual Learning (FCL) is fundamental to real-world distributed learning systems, requiring models to adapt to sequential, non-IID data across clients while mitigating catastrophic forgetting and client drift. Existing approaches formulate ...

📖 Read original article


12. Beyond Sparse Weights: When Is Attention Compressible? ​

Author: Chiwun Yang, Xiaoyu Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.21541v1 Announce Type: new Abstract: KV-cache compression is often justified by attention maps with a few large weights. This is incomplete: large weights may not contain most of the mass, omitted values can cancel, and preserving the attention output may not preserve the task. We separat...

📖 Read original article


13. Loss-Parameterized Fisher Width Along Learning Trajectories ​

Author: Vu Khac Ky
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.TH

arXiv:2608.21561v1 Announce Type: new Abstract: Fisher width measures the Gaussian width of a probe set after deformation by the local Fisher geometry. We study its evolution along learning trajectories and ask when training loss can serve as an effective coordinate for this quantity. We first deriv...

📖 Read original article


14. Anchoring Bias: A Persistent Fairness Backdoor Attack against MLLMs under Continual Learning ​

Author: Yuyang Luo, Kai Shu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21577v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are increasingly deployed in high-stakes domains where fairness is a critical safety requirement. In practice, these models are continually updated through continual learning (CL) to adapt to evolving tasks and ...

📖 Read original article


15. Sorting from Counterexamples ​

Author: Noga Alon, Shay Moran, Shlomo Moran
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CC, cs.CG, cs.DS, math.CO

arXiv:2608.21579v1 Announce Type: new Abstract: Consider the following problem of learning an unknown linear order on $n$ items. In each round, the learner guesses a complete ordering of the items and receives either confirmation that the guess is correct or a counterexample: a pair of items in the ...

📖 Read original article


16. Reading the Room: Implicit Confusion Encoding in Recurrent World Model States ​

Author: Donald Aadithiyan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21582v1 Announce Type: new Abstract: World models built on the RSSM architecture, such as DreamerV3, keep a recurrent hidden state $h_t$ trained only to reduce prediction error. We show this state also tracks its own confusion, hiding in plain sight: nearly orthogonal to $h_t$'s direction...

📖 Read original article


17. Predicting Early Functional Decline from Longitudinal Laboratory and Vital Sign Trajectories: A Large-Scale Study Using the All of Us Research Program ​

Author: Rashmita Kudamala, Aravind V. Kuruvikkattil, Lalitha Pranathi Pulavarthy, Saptarshi Purkayastha
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, stat.AP

arXiv:2608.21589v1 Announce Type: new Abstract: Functional decline in older adults is typically recognized only after falls or observable gait impairment, closing the window for prevention. We investigated whether temporal trajectories of routine biomarkers, already recorded but rarely analyzed long...

📖 Read original article


18. Reaching the Tail: Calibration Diversity Drives Conformal Coverage under Data Scarcity ​

Author: Donald Aadithiyan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21591v1 Announce Type: new Abstract: Multi-horizon rare-event forecasting is hard under long macroeconomic series' data constraints: labeled events are scarce, and standard uncertainty quantification assumes an exchangeability that autocorrelation violates. A controlled ablation shows an ...

📖 Read original article


19. Subzero matrix completion for sparse data analysis: large-scale learning of latent low-rank structure ​

Author: Lawrence K. Saul, Ningyuan Huang, Dennis Bollweg, Jeff Soules, Diana C. Halikias
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.21607v1 Announce Type: new Abstract: We investigate when a sparse nonnegative matrix can be recovered from a real-valued matrix of much lower rank by zeroing out its negative elements. The potential for such decompositions suggests a mathematical connection between sparsity and rank; we a...

📖 Read original article


20. FrugalSOT - Frugal Search Over the Models ​

Author: Pradheep P, Yuvanesh S, Harish KB, Keerthan Saai Reddy S, Joshva Devadas T, Naveenkumar J, Hemalatha K
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21621v1 Announce Type: new Abstract: In on-device NLP tasks, limited resources of embedded hardware, such as the Raspberry Pi 5, require efficient inference strategies. This paper introduces FrugalSOT (Frugal Search Over The Models), a resource-aware model selection architecture for on-de...

📖 Read original article


21. Rethinking Communication Metrics: How Should We Measure Meaning? ​

Author: Niloofar Tavakolian, Hakimeh Purmehdi, Jungyeon Baek
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21626v1 Announce Type: new Abstract: Semantic communication shifts the objective of communication systems from accurate symbol reconstruction toward meaning preservation, task accomplishment, and efficient information exchange. However, its evaluation remains fragmented across telecommuni...

📖 Read original article


22. ChequeMark: An Ensemble Machine Learning Framework for After-Hours Business Deposit Fraud Detection ​

Author: Ann Youduo Xu, Emily Yu, Justin Leski, William Lam
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21629v1 Announce Type: new Abstract: Cheque fraud is a material risk in after-hours business deposit operations because funds may be released within one business day, while cheque clearing takes several days. This timing gap creates a fraud exposure window for financial institutions. Prio...

📖 Read original article


23. Large-Scale Evaluation of Advanced Imputation Methods for Missing Values in Smart Meter Data ​

Author: Daniela Stojcheska, Marija Markovska, Dimitar Taskovski, Branislav Gerazov, Boris Nikolov
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21638v1 Announce Type: new Abstract: Accurate and reliable collection of electricity consumption data through Advanced Metering Infrastructure (AMI) is of great importance for the operation of smart grids, especially for the detection of non-technical losses (NTL). However, real-world dat...

📖 Read original article


24. Power-Performance Characterization of TinyML Systems ​

Author: Yujie Zhang, Dhananjaya Wijerathne, Zhaoying Li, Tulika Mitra
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21646v1 Announce Type: new Abstract: TinyML systems are enabling machine learning (ML) inference at the edge. However, there is little quantitative analysis of such systems. This paper presents a systematic performance and power characterization of diverse TinyML applications on microcont...

📖 Read original article


25. GeoQ: Geometry-Aware Conditional Quantile Error Estimation for Scientific Surrogate Models ​

Author: Khoa Nguyen, Daniel Serino, Aviral Prakash, Marc Klasky
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.21652v1 Announce Type: new Abstract: Neural-network surrogate models are increasingly used to accelerate scientific simulations, but their deployment in extrapolative and autoregressive settings requires input-dependent estimates of prediction error. In this work, we introduce GeoQ (Geome...

📖 Read original article


26. Bounded Precision-Geometry Scaling for Robust Multi-Task Learning under Loss Scale Mismatch ​

Author: Krishna Subedi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.21653v1 Announce Type: new Abstract: Multi-task learning often combines losses that span several orders of magnitude, causing homoscedastic uncertainty weighting to degrade severely. We propose Bounded Precision-Geometry Scaling (BPGS), a method that maps each task's log-variance through ...

📖 Read original article


27. Variational Structure at the Edge of Stability ​

Author: Eric Regis
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21660v1 Announce Type: new Abstract: When discrete-time optimizers operate at the edge of stability, they exhibit near-two-periodic behavior. These oscillatory dynamics are reminiscent of conservative systems, such as the dynamics generated by symplectic integrators. However, a precise fo...

📖 Read original article


28. SynEHR: Joint Modeling Inter-visit Temporal Evolution and Intra-visit Clinical Structure for Longitudinal EHR Synthesis ​

Author: Ximiao Li, Lin Jiang, Rongchao Xu, Dahai Yu, Zhe He, Guang Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21673v1 Announce Type: new Abstract: Longitudinal electronic health records (EHRs) document patients' sequences of clinical visits over time, preserving the temporal evolution of disease progression and care delivery. However, real longitudinal EHRs are difficult to access because they co...

📖 Read original article


29. Read, Write, Relax: Why Neural PDE Surrogates Need Both Global and Local Processing ​

Author: Anuj Kumar, Heiko Zimmermann, Josiah Bjorgaard, Jacan Chaplais, Nikolaos Bouklas, Matteo Salvador, Alexander Lavin
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE, physics.app-ph, physics.comp-ph

arXiv:2608.21677v1 Announce Type: new Abstract: Recent mesh-based simulation advances have, in no small part, relied on neural surrogates of two distinct families: global models that route information through a small set of latent tokens, and local models that perform message passing across mesh edg...

📖 Read original article


30. Benchmarking Composable Compression Techniques in Mixture-of-Experts LLMs ​

Author: Afsara Benazir, Chen Chen, Rongxiao Qu, Jiabo Huang, Jingtao Li, Lingjuan Lyu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21693v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) LLMs scale model capacity efficiently through sparse activation, but their large expert parameter footprint, routing imbalance, and long-context KV-cache growth make deployment difficult on commodity hardware. Practical deploym...

📖 Read original article


31. Posterior Information Dynamics of Diffusion Models for Linear Inverse Problems ​

Author: Xiangming Meng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, eess.SP, math.IT

arXiv:2608.21709v1 Announce Type: new Abstract: Diffusion models are widely used as priors for linear inverse problems, yet endpoint quality does not reveal when measurement information enters reverse denoising or how it is allocated across signal directions. We study this process through the smooth...

📖 Read original article


32. Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data ​

Author: Renfei Zhang, Niloofar Mireshghallah
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21727v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is deployed to make models better at reasoning tasks, but its side effect on what models will divulge is under studied. Here we show that RLVR on facts increases extraction of personally identifiabl...

📖 Read original article


33. Adaptive Multilevel Twisted Sequential Monte Carlo for Rare Events Estimation in Language Models ​

Author: Zixuan Liu, Fangzheng Wu, Brian Summa, Zizhan Zheng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21736v1 Announce Type: new Abstract: Rare unsafe behaviors in large language models can remain practically significant even when their probability is extremely small, particularly at deployment scales involving millions or billions of interactions. Twisted Sequential Monte Carlo (SMC) pro...

📖 Read original article


34. Width-Independent Compressibility of Deep Neural Networks ​

Author: Hong-Yi Wang, Mingze Wang, Liu Ziyin
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, cs.IT, math.IT

arXiv:2608.21752v1 Announce Type: new Abstract: It has long been known that well-trained neural networks can be compressed very strongly without affecting their performance, an important phenomenon that remains poorly understood. We prove a uniform compressibility theorem for deep multilayer percept...

📖 Read original article


35. How Architecture and Training Affect TPC Representations Across Experiments ​

Author: Tyler Wheeler, Michelle P. Kuchera, Raghuram Ramanujan, William Sieland, Ryan Krupp, Daniel Bazin, Connor L. Cross, Hoi Yan Ian Heung, Andrew J. Jones, Ruchi Mahajan, Saiprasad Ravishankar, Pranjal Singh, Benjamin Votaw, Chris Wrede
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, nucl-ex, physics.ins-det

arXiv:2608.21756v1 Announce Type: new Abstract: Deep-learning efforts have increasingly shifted toward foundation model approaches. In experimental physics, this allows models and learned representations to be reused beyond the experiments in which they were developed. This work evaluates the reusab...

📖 Read original article


36. A Fixed-Radius Distance-Band Benchmark for Dimensionality-Reduction Fidelity ​

Author: Yoshio Takaeda
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21779v1 Announce Type: new Abstract: Dimensionality-reduction (DR) methods are routinely judged by how well each point's k nearest neighbors survive the 2-D embedding (recall@k, trustworthiness, continuity). We argue this family is a biased measure of distance fidelity: its per-point vari...

📖 Read original article


37. MSM-Mem: A Universal Medical Structured Multimodal Memory Framework for Medical AI Agents ​

Author: Md Asaduzzaman Jabin, Khoa Le, Lin Zhao, Tianming Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.21810v1 Announce Type: new Abstract: Clinical decision-making is inherently experience-driven: physicians progressively refine their reasoning by synthesizing patient history, multimodal observations, and prior diagnostic experiences across interactions. In contrast, current multimodal la...

📖 Read original article


38. Resilient Concurrent Causal Discovery for Topological Event Sequences ​

Author: Jiyu Tian, Junhao Dong, Mingchu Li, Lingling Fang, Liming Chen, Andreas Holzinger, Zheng Yan, Yew Soon Ong
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21815v1 Announce Type: new Abstract: Causal discovery on topological event sequences is crucial for ensuring the reliability of networks. However, existing methods struggle to capture the complex causal relationships arising from concurrent events and lack robustness to incomplete event s...

📖 Read original article


39. A Physics-informed Neural Network Approach for Robust Buckling Load Prediction and Reliability-Based Design of Thin Truncated Conical Shells ​

Author: Devasmit Dutta, Budhaditya De, Rohan Majumder, Sudip Kumar Mishra
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21818v1 Announce Type: new Abstract: Thin-walled truncated conical shells are widely used in aerospace, marine, offshore, and lightweight infrastructure systems due to their high strength-to-weight ratio and geometric efficiency. Their buckling resistance under axial compression, however,...

📖 Read original article


40. More Experts, Worse Dynamics: Inverse Scaling and Spectral Bias in Mixture-of-Experts State-Space Models ​

Author: Chandresh Pandey
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21840v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures are commonly motivated as a way to increase expressivity by decomposing complex systems into simpler local dynamics. This intuition has recently been extended to spectral state-space models, where mixing stable op...

📖 Read original article


41. ChainPrune: Evaluating and Reducing Redundancy in Long Chain-of-Thought Reasoning ​

Author: Weihang Pan, Zhengxu Yu, Yuxiang Zhang, Wenzhi Li, Zhongming Jin, Binbin Lin, Xiaofei He, Jieping Ye
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21860v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning has significantly enhanced the multi-step problem-solving capabilities of large language models (LLMs) by introducing explicit intermediate reasoning. However, advanced Large Reasoning Models (LRMs) often exhibit overth...

📖 Read original article


42. BioMed-Agent-RL: A Meta Learning, All You Need for Biomedical Applications ​

Author: Md Asaduzzaman Jabin, Zihao Wu, Tianming Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.21864v1 Announce Type: new Abstract: The current progress of Clinical Vision Large Language Models (C-VLLMs) has substantially improved digital diagnostics, still these frameworks often endure lesion noises, modality misalignment, hallucination, and missed contextual grounding in complex ...

📖 Read original article


43. CD-LoRA: Consistency-Driven Low-Rank Adaptation for Multi-Task Fine-Tuning ​

Author: Qian Zha, Jinda Liu, Yuan Wu, Yi Chang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21909v1 Announce Type: new Abstract: While Multi-Task Learning (MTL) is essential for adapting Large Language Models (LLMs) to diverse domains, prevailing LoRA-based methods rely on complex routing mechanisms that partition task-specific knowledge. In this work, we reveal that such routin...

📖 Read original article


44. Beyond Fixed Directions: Adaptive Representation Analysis of Reasoning and Memorization in LLMs ​

Author: Shaheen Nabi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21919v1 Announce Type: new Abstract: Recent work has proposed that reasoning and memorization in language models can be characterized by a single representation direction, including methods that keep this direction fixed during reinforcement learning. We test two assumptions behind this v...

📖 Read original article


45. Bi-EZP: LLM-Guided Bilevel Program Evolution for Ensemble Zero-Cost Proxy Discovery ​

Author: Yutao Lai, Kezhao Lai, Hai-Lin Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21927v1 Announce Type: new Abstract: Zero-cost proxies enable neural architecture search (NAS) to rank candidate networks from statistics computed at initialization, avoiding repeated training. However, different proxies capture different properties and often produce inconsistent rankings...

📖 Read original article


46. Gated Decoupled Compositional Bandits: A Unified Theory of Contextual Bandits with Supervised-Calibrated Action Scaling and Pre-Execution Gating ​

Author: Oleg Miroshnichenko
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.21993v1 Announce Type: new Abstract: We introduce Gated Decoupled Compositional Bandits (GDCB), a family of contextual bandit algorithms with three structural innovations that jointly fall outside the taxonomy of LinUCB, LinTS, HierTS, factored bandits, neural contextual bandits, and RLHF...

📖 Read original article


47. Variance Driven Exploration: A Provable and Efficient Methodology for Pure Exploration in Highly Stochastic Environments ​

Author: Khang Luong, Nam Nguyen, Hoang Ta, Hung The Tran, Tuan Dam
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.21995v1 Announce Type: new Abstract: We propose Variance Driven Exploration (VarDE), a principled approach for pure exploration in highly stochastic environments, where the exploration process is dominated by stochastic variance. VarDE is built on a fundamental principle: sampling effort ...

📖 Read original article


48. DySCo: Dynamically consistent data-driven downscaling of extremes in climate projections ​

Author: S. Stamatelopoulos, M. Wang, I. Lopez-Gomez, L. Zepeda-Nunez, Z. Y. Wan, R. Carver, F. Sha, T. P. Sapsis
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, physics.ao-ph

arXiv:2608.21998v1 Announce Type: new Abstract: Regional climate risk assessment is critical for applications such as infrastructure design, disaster forecasting, and insurance resource allocation. However, estimating regional (i.e., high-spatial-resolution) risk with global climate models (GCMs) re...

📖 Read original article


49. The Communication Map of a Transformer ​

Author: Richard Zhe Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.22007v1 Announce Type: new Abstract: The components of a transformer communicate by writing to and reading from a shared residual stream, and mechanistic interpretability has mapped these connections by hand, one circuit at a time. We present the communication map, which charts every pote...

📖 Read original article


50. Spectral Pre-Filtering for Context-Adaptive Sensor Fusion: A Four-Role FFT-GDCB Integration for High-Stakes Decision Systems ​

Author: Oleg Miroshnichenko
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22023v1 Announce Type: new Abstract: Context-adaptive Kalman filters calibrate their noise covariance matrices Q and R from innovation residuals via online regression. When the underlying sensor or signal carries periodic structure -- mechanical LiDAR rotation harmonics, engine vibration,...

📖 Read original article


51. ARCHER: Amortized cross-specimen pose estimation for cryo-electron microscopy ​

Author: Nhan D. Nguyen, Bao Pham
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, q-bio.BM, q-bio.QM

arXiv:2608.22029v1 Announce Type: new Abstract: Single-particle cryo-electron microscopy (cryo-EM) pose estimation is traditionally solved anew for each dataset, where iterative refinement is done from scratch while the estimator learns to store the molecule in its weights. In this work, we show tha...

📖 Read original article


52. ReMAP: Self-supervised learning to unveil brain representations and vulnerability ​

Author: Jade Perdereau, Virginie Loison, Kanssa El Ayeb, Louis Gervais, Melvin Berto Strouc, Fabrice Vall'ee, Thomas Moreau, J'er^ome Cartailler
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22042v1 Announce Type: new Abstract: General anesthesia offers a rare opportunity to observe the human brain under a standardized, controlled perturbation. Yet intraoperative electroencephalography (EEG) is almost always reduced to a single proprietary depth index, collapsing a rich traje...

📖 Read original article


53. Personalized and Aspiration-Oriented Career Path Recommendation ​

Author: Kuleshwar Sahu, Girish Keshav Palshikar, Rajiv Srivastava
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22056v1 Announce Type: new Abstract: Fulfilling career aspirations is important for growth of employee and organization. We propose a data driven methodology to recommend personalized career path for a given aspirant's career path and aspirations. The pro-posed method uses the career path...

📖 Read original article


54. Improving Energy Efficiency of Oil Platforms Through Optimal Loading of Diesel Generators Using Machine Learning and Search Algorithms ​

Author: Khivishta Boodhoo, Josh Plumbly, Nicholas Watson
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22076v1 Announce Type: new Abstract: Rising energy demand, fossil fuel depletion and climate change highlight the need for more efficient energy production and consumption. Offshore oil and gas platforms face challenges related to inefficient energy use, system failures, accessibility and...

📖 Read original article


55. Counterfactual Quotient Models: Learning What Actions Change, Not What the World Does ​

Author: Junlin Chen, Ruijie Wang, Jianxin Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22092v1 Announce Type: new Abstract: Reinforcement-learning models commonly predict complete future states, observations, or feature occupancies, even though action selection depends only on differences between the consequences of candidate actions. As a result, these models may devote su...

📖 Read original article


56. Beyond Fresh Starts: Stateful Inference for Streaming ASR in Conversational Voice Agents ​

Author: Sameep Chattopadhyay, Alexander Erdmann, Mari Ostendorf
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22101v1 Announce Type: new Abstract: Modern voice-agent systems rely on streaming speech recognition models that operate under stringent latency constraints. This study shows that, due to the limited memory constraints of real-time processing, these systems are adversely impacted by conve...

📖 Read original article


57. What actually runs: a measurement study of language model placement and decode speed on the Apple Neural Engine ​

Author: Shahir M A
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.PF

arXiv:2608.22110v1 Announce Type: new Abstract: We ask what gets a language model onto the Apple Neural Engine (ANE) and what makes it fast there, and we answer with three measurements. We sweep a 64-shape matrix of LLM primitives that varies how a computation is expressed while holding what it comp...

📖 Read original article


58. Symbolic Neural ODEs: Learning interpretable models from time-series data ​

Author: Nibodh Boddupalli, Jeff Moehlis
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY, math.DS, stat.ML

arXiv:2608.22112v1 Announce Type: new Abstract: We present a machine learning framework for identifying sparse, interpretable models of dynamical systems directly from time-series data. Our approach parameterizes the underlying vector field using a neural architecture and trains it by minimizing a m...

📖 Read original article


59. CST: Collaborative Selective Transmission for Communication-Efficient Multimodal Edge Inference ​

Author: Hai Chi, Junrui Zhang, Rui Ning, Chonggang Wang, Robert Gazda, Huanrui Yang, Hongyi Wu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22115v1 Announce Type: new Abstract: Collaborative multimodal inference improves edge perception by combining observations from distributed sensing devices, but transmitting high-dimensional helper representations incurs substantial communication overhead and can lead to high end-to-end l...

📖 Read original article


60. TANGO: Token-Aggregated Nonlinear Gating Operators for Natural and Formal Language Modeling ​

Author: Joshua Nunley
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.22117v1 Announce Type: new Abstract: A standard Transformer block separates cross-token interaction in self-attention from a nonlinear feed-forward network applied independently at each position. We introduce the TANGO model (Token-Aggregated Nonlinear Gating Operators), which replaces th...

📖 Read original article


61. The Price of Decentralization in Top-$K$ Arm Identification ​

Author: Larissa Xu, Jasmine Nguyen, William Chang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22120v1 Announce Type: new Abstract: Cooperative teams often need to agree on the best few options rather than simply accumulate reward, and they must do so while each member sees only a fragment of the team's collective experience. We study this as top-$K$ joint-arm identification in mul...

📖 Read original article


62. Decoupled Physical Modeling and Execution for Physics Reasoning ​

Author: Ye Zhang, Xuehang Guo, Rui Pan, Pengfei Yu, Denghui Zhang, Manling Li, Qingyun Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.22126v1 Announce Type: new Abstract: Physics reasoning requires constructing a consistent model of the underlying physical system rather than relying solely on symbolic or formula-based manipulation. Although large language models have shown strong ability in solving math and coding probl...

📖 Read original article


63. Who Should Teach? Confidence-Aware Dual-Teacher Learning for Few-Shot Node Classification on Text-Attributed Graphs ​

Author: Hojin Kim, Sujin Yoon, Sungsu Lim, Dongwon Lee, David Yoon Suk Kang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.SI

arXiv:2608.22127v1 Announce Type: new Abstract: Text-Attributed Graphs (TAGs) integrate graph structures and node-associated textual attributes, and recent studies have increasingly leveraged Large Language Models (LLMs) to improve TAG learning in few-shot settings. However, existing approaches typi...

📖 Read original article


64. Blockwise Stabilized Adaptive Cubic Regularization with Subsolvers via Recurrence ​

Author: Rodion Podorozhny
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2608.22129v1 Announce Type: new Abstract: Cubic-regularized Newton methods have the optimal $\mathcal{O}(\varepsilon^{-3/2})$ global rate and an automatic saddle-escape mechanism, but their subproblem is most often solved by a full eigendecomposition, limiting feasible model size. We introduce...

📖 Read original article


65. Learning Reduced-Order Dynamics with Singularity via Latent-Augmented Neural Ordinary Differential Equations ​

Author: Xiaorui Wang, Yu Zhou, Wenjie Mei, Dongzhe Zheng, Yang Bai, Masaaki Nagahara
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22142v1 Announce Type: new Abstract: This paper addresses the issue of self-intersecting trajectories (in phase space) in industrial reduced-order modeling and proposes the Latent-Augmented Neural Ordinary Differential Equations (LA-NODEs) framework. From the perspective of artificial int...

📖 Read original article


66. Loss Landscape Features That Make Adam Stall: Definitions, Estimators, and the Preconditioned Hessian View ​

Author: Rodion Podorozhny
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2608.22145v1 Announce Type: new Abstract: Across implicit-neural-representation (INR) architectures and analytic benchmarks we observe that a thoroughly tuned Adam (especially its learning rate (lr), e.g. in a hyperparameter sweep from $lr = 0.05$ to $10^{-8}$) can potentially reach a very low...

📖 Read original article


67. More accurate behavioral predictions with hybrid Bayesian-connectionist models ​

Author: Brenden M. Lake, Akshay K. Jagadish, Guangyuan Jiang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22154v1 Announce Type: new Abstract: Researchers must often choose between Bayesian or neural network models of behavior, two paradigms with complementary strengths and weaknesses. An ideal paradigm would facilitate testing many kinds of representations and inductive biases; Bayesian mode...

📖 Read original article


68. Why Does Robustness Reduce Superposition? ​

Author: Adam Elimadi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22155v1 Announce Type: new Abstract: The study of adversarial examples and their origins remains an open area of research. Mechanistic interpretability, and superposition in particular, offers new avenues for approaching this problem. Gorton & Lewis (2025) demonstrate that adversarial exa...

📖 Read original article


69. Unveiling the Depth-Performance Dilemma in Split-Federated Fine-tuning of LLMs ​

Author: Hariharan Ramesh, Someshwaran Murugaiyan, Jyotikrishna Dass
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.DC

arXiv:2608.22188v1 Announce Type: new Abstract: Split Federated Fine-tuning (SFF) is a promising paradigm for scaling Large Language Models (LLMs) by partitioning model depth between resource-constrained clients and a centralized server. While system incentives for throughput and privacy favor deep ...

📖 Read original article


70. On the Capability Separation Between World-Model Policy Learning and Imitated World-Action Models ​

Author: Yang Yu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22197v1 Announce Type: new Abstract: World-action models predict a future outcome and then infer an associated action. Although this factorization can improve representation learning and data efficiency, it is unclear whether it provides stronger control capability than direct behavior cl...

📖 Read original article


71. Joint Causal Structure and Cluster Discovery Using Variational Inference ​

Author: Avni Rajpal, Anubhav Kumar, Rishabh Karnad, Mohammad Emtiyaz Khan, P. K. Srijith
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.22212v1 Announce Type: new Abstract: Causal discovery aims to understand the relationships between individual random variables. In many applications, such as brain imaging and climate modeling, it is more meaningful to consider interactions among groups of variables. Existing methods assu...

📖 Read original article


72. Counterfactual Evaluation of Temporal Observation Protocols ​

Author: Xizhe Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22221v1 Announce Type: new Abstract: We study counterfactual protocol evaluation: whether data collected under a realised observation protocol determine the predictive value of alternatives that were never deployed. Protocol value is the population $R^2$ of the Bayes-optimal predictor of ...

📖 Read original article


73. FreKoo++: Learning Continuous Spectral Dynamics for Temporal Domain Generalization ​

Author: En Yu, Xiaoyu Yang, Wei Duan, Guangquan Zhang, Jie Lu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22224v1 Announce Type: new Abstract: Temporal Domain Generalization (TDG) aims to learn from historical domains and generalize to unseen future distributions under concept drift. Nevertheless, prevailing TDG methods struggle with complex real-world streaming scenarios involving both multi...

📖 Read original article


74. Risk-Sensitive Reinforcement Learning with Smoothed Quantile Objectives ​

Author: Mohammad Alipour-Vaezi, Huaiyang Zhong, Sajad Khodadadian
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.22227v1 Announce Type: new Abstract: Reinforcement Learning (RL) has achieved tremendous success in recent years. However, the classical foundations of RL do not account for the risk sensitivity of the objective function, which is critical in various fields, including healthcare, finance,...

📖 Read original article


75. When Test-Time Adaptation Helps, Harms, or Becomes Inactive: A Condition-Level Study on CIFAR-10-C ​

Author: Sreeja Guha Majumdar, Aratrika Saha
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.22233v1 Announce Type: new Abstract: Test-time adaptation (TTA) aims to improve model robustness under distribution shift by adapting a source model using unlabeled test data. Although methods such as TENT and EATA have demonstrated gains on corrupted data, aggregate accuracy can obscure ...

📖 Read original article


76. A Query-Time Framework for Transient 2D Pore-Scale Flow Prediction and Generative Design ​

Author: Yiming Wang, Jiale Zhu, Zhichen Ye, Yandong Lv, Shiqi Wang, Jinlong Liu, Yucheng Fan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22235v1 Announce Type: new Abstract: Pore-scale flow governs transport and permeability behaviour in porous media engineering applications, yet repeated lattice Boltzmann method (LBM) simulation across many geometries and design queries remains costly for repeated deployment. This study f...

📖 Read original article


77. Toward a First-Principles Update Geometry for the Language-Model Head ​

Author: Aditya Somasundaram
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22253v1 Announce Type: new Abstract: We study the language-model head and softmax as a single module, deriving an update geometry from their composition rather than from the weight matrix in isolation. Under Hilbert's projective distance, the maximum change caused by an update $S$ over $...

📖 Read original article


78. DAW: Dynamics-Aware Weighting for Deep Learning Forecasts of Chaotic Systems ​

Author: Zhou Fang, Gianmarco Mengaldo
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph

arXiv:2608.22277v1 Announce Type: new Abstract: Deep learning surrogates for forecasting chaotic dynamical systems suffer from catastrophic error accumulation over long-term autoregressive rollouts. This behavior is partly tied to the underlying systems: chaotic spatiotemporal systems, such as the K...

📖 Read original article


79. StocBench: A Benchmark for Generative Modeling of Stochastic Dynamics ​

Author: Sebastian Pfister, Benjamin Holzschuh, Nils Thuerey
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22309v1 Announce Type: new Abstract: We benchmark transport-based generative models as well as distillation-based few-step methods for the probabilistic forecasting of stochastic fluid flows, with a particular focus on performance under limited inference budgets. All methods are evaluated...

📖 Read original article


80. Does a Modern-Handwriting Warm-Up Help Historical Arabic OCR? A Reproducible, Compute-Matched Evaluation on Muharaf and KHATT ​

Author: Sumaih Almarshad, Maram Alamri, Dona Aloraini, Fares Altuwaim, AlJawharh AlOtaibi, Reem Alyabis, Rayah Aldawsari
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.22316v1 Announce Type: new Abstract: Whether an intermediate stage of modern Arabic handwriting helps or hurts historical Arabic HTR is usually decided from one implementation and one comparison, too thin a basis for a claim either way. We test stability by running the same nominal ablati...

📖 Read original article


81. Beyond Dense Adam States: Adaptive Log-Space Quantization for Memory-Efficient Optimizers ​

Author: Yan Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22322v1 Announce Type: new Abstract: Low-precision optimizer-state methods are commonly designed for dense Adam-style moments, but memory-efficient optimizers maintain factored, confidence-based, or projected states whose quantization errors propagate differently. We characterize this het...

📖 Read original article


82. Gaussian process learning with flow map refinement for parameter estimation in dynamical systems ​

Author: Yue Hao, Dongwei Ye
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.NA, math.NA

arXiv:2608.22324v1 Announce Type: new Abstract: Parameter estimation is a central task in data-driven learning of dynamical systems. It aims to recover the underlying physical parameters from observed time-series data, thereby providing interpretable insights into the physical mechanisms governing t...

📖 Read original article


83. SANE: State Anomaly Neutralization for Stable Extreme-Context Delta-Rule Models ​

Author: Qingwen Lin, Boyan Xu, Xiao Liu, Zhifeng Hao, Ruichu Cai
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22354v1 Announce Type: new Abstract: Delta-Rule recurrent models maintain a fixed-size state, enabling $O(1)$ inference memory but potentially becoming unstable under extreme-context extrapolation. By tracking RWKV-7 over sequences of up to 100M tokens, we empirically identify a distinct ...

📖 Read original article


84. Tracing the Unlabeled Storm: Cross-Variable Transfer in a Lagrangian Atmospheric JEPA Framework ​

Author: K M Anirudh, S Sandeep, Hariprasad Kodamana
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, physics.geo-ph

arXiv:2608.22358v1 Announce Type: new Abstract: Deep atmospheric convection governs South Asian monsoon variability, yet attempting to learn its latent world model directly from zero-inflated, heavy-tailed precipitation yields suboptimal predictive representations. Continuous atmospheric proxies, su...

📖 Read original article


85. Self-Supervised Graph Representation Learning for In-The-Wild Wearable and Smartphone based Emotion Recognition ​

Author: Ioannis N. Ziogas, Leontios J. Hadjileontiadis, Ahsan H. Khandoker, Aamna Al Shehhi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP

arXiv:2608.22387v1 Announce Type: new Abstract: Wearable and smartphone-based emotion recognition (WER) remains a challenging setting in affective computing, due to the notorious difficulty and bias associated with in-the-wild label collection. The high inter-and intra-subject emotional variability ...

📖 Read original article


86. Dual-Scale State-Space Modeling with Speaker-Wise Dynamic CRF for Speech Emotion Recognition in Conversation ​

Author: Guan-Hua Wen, Kuan-Yu Chen, Hou-Chiang Tseng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22399v1 Announce Type: new Abstract: Conversational speech emotion recognition must reconcile acoustic evidence across temporal scales with two interaction processes: cross-speaker contextual influence and within-speaker emotion evolution. We propose DSSM-CRF, an audio-only architecture t...

📖 Read original article


87. Geometric Structures on Graphs: a Holonomy-Based Discretization of Curvature ​

Author: Hao Li, Yuhan Peng, Junwen Dong
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, math.DG

arXiv:2608.22453v1 Announce Type: new Abstract: We propose a holonomy-based framework for discretizing curvature on graphs equipped with local symmetric positive-definite metrics. Each vertex carries a fibre metric (g_i), and each directed edge carries a reversible metric-compatible transport (F_...

📖 Read original article


88. KPI-Conditioned Generative Design of Automotive Hood Inner Panels: A Two-Stage Retrieval-Generation Pipeline with Surrogate-Based Performance Estimation ​

Author: Sudeep Chavare
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22457v1 Announce Type: new Abstract: An inner hood panel must meet a deflection target, stay below a stress limit, and hit a mass target. Machine-learned surrogates have made the forward direction, geometry to performance, fast and routine. The inverse direction, producing geometry from a...

📖 Read original article


89. MASH-Bench: Diagnosing Cross-Source Failure in Mass-Shooting Risk Classification ​

Author: Neha Sharma, Ritesh Sharma
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2608.22460v1 Announce Type: new Abstract: Public mass-shooting databases differ substantially in coverage, feature availability, and reporting practices, creating challenges for machine-learning models that must generalize across data sources. We introduce MASH-Bench, a harmonized benchmark of...

📖 Read original article


90. Functional compatibility as a determinant of persistent neural learning ​

Author: Hossein Javidnia
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22462v1 Announce Type: new Abstract: Artificial neural networks can acquire new capabilities but often damage existing ones when they continue to learn. This stability-plasticity problem has motivated replay, regularization and constrained-update methods, yet it remains unclear whether a ...

📖 Read original article


91. The Variance of Thought: Policy Variance, Critical Forks, and Local Credit Assignment ​

Author: Yingru Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22467v1 Announce Type: new Abstract: Long-horizon language-model tasks --- multi-step reasoning and tool-using agents alike --- are limited by credit assignment. We analyze it through the policy variance $\sigma_\pi^2(s)=\operatorname{Var}{a\sim\pi}[Q{\pi}(s,a)]$, which in a determinist...

📖 Read original article


92. Quantum-Inspired Hybrid Neural Networks for Neural Decoding: A Controlled Ablation Study of Learnable Quantum Sidecar Integration ​

Author: Diana Legziel Levy, Menachem Finkelstein, Peter Chin, Eilon Vaadia, Sarel Cohen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, quant-ph

arXiv:2608.22475v1 Announce Type: new Abstract: We study parameterized quantum circuits (PQCs) integrated as residual sidecar modules within a ResNet-50 backbone for 31-class neural population decoding---imagined handwriting classification from multi-neuron spike rasters. Under strictly controlled c...

📖 Read original article


93. From Symmetry to Invariance: Learning Galois Equivalent Representations in Finite Fields ​

Author: Zheng Zhang, Na Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22513v1 Announce Type: new Abstract: Neural networks can learn algebraic operations from finite examples, but it remains unclear whether this ability transfers across mathematically equivalent representations of the same operation. We study this question through multiplication in finite f...

📖 Read original article


94. From Detrimental to Beneficial: Dynamic Influence-based Valuation and Editing ​

Author: Adrian Nyakairu, Hongfu Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22522v1 Announce Type: new Abstract: Data valuation is a cornerstone of data-centric learning, where prior efforts primarily focus on designing algorithms to classify training samples as either beneficial or detrimental for the learning task. However, leveraging these valuation estimates ...

📖 Read original article


95. Stress Testing Unlearning Algorithms ​

Author: Noam Diamant, Ethan Fetaya, Neta Glazer
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22527v1 Announce Type: new Abstract: Recently, machine unlearning, the removal of specific training data influence from a model, has gained increasing attention. In large language models (LLMs), unlearning is particularly challenging due to the ambiguity of inputs and outputs. Con- sequen...

📖 Read original article


96. BLADE: Bilevel Low-rank Augmented-Lagrangian Erasure for LLM Unlearning ​

Author: Md Toufikuzzaman, Ahmad Mousavi, Dongwon Lee
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.22557v1 Announce Type: new Abstract: Existing LLM unlearning methods struggle with robustness: unbounded forget losses degrade model coherence, fixed-weight balancing cannot adapt as retain difficulty shifts mid-training, and methods that work on one benchmark falter under scaling or repe...

📖 Read original article


97. Clinical Graph-JEPA: Predictive Patient-State Knowledge Graphs for Cognitive Decision Support ​

Author: Kushagra Yadav, Nalin Prabhath, Amit Lamba, Goeun Han, Yining Mao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22583v1 Announce Type: new Abstract: Clinical records contain rich evidence about patient state, but converting that evidence into reliable, structured knowledge graphs remains difficult because extraction errors, ontology mismatch, missing relations, and temporal ambiguity can propagate ...

📖 Read original article


98. GCA: Global Centroid Alignment in Federated Learning ​

Author: Jong-Ik Park, Harry Jiang, Logan Blakely, Georgios Fragkos, Shamina Hossain-McKenzie, Carlee Joe-Wong
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2608.22593v1 Announce Type: new Abstract: Autoencoder (AE)-based federated learning (FL) is attractive for anomaly detection when clients have limited local data. However, conventional FL exchanges AE parameters or gradients, incurring substantial communication overhead and potentially exposin...

📖 Read original article


99. Tabular foundation models for non-tabular tasks ​

Author: Goran Nakerst, John Brennan, Wouter Beugeling, Masudul Haque
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22594v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have recently emerged as a promising paradigm for machine learning on tabular data, offering the ability to generalize across datasets without task-specific training. Since many machine learning datasets can be represen...

📖 Read original article


100. Adversarial Agents on Topology Optimization: Understanding the Fragility and Robustness of Deep Learning-based and Physics-Based Design Models under Adversarial Perturbation ​

Author: Hoang Anh Nguyen, Yuan Hong, Hongyi Xu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22606v1 Announce Type: new Abstract: Topology optimization, using both physic-based approaches and deep learning surrogates, serves as a cornerstone for generative design agents in cyber-manufacturing systems. While deep learning surrogates have gained widespread adoption due to their spe...

📖 Read original article


101. Mitigating Explanation Leakage in Financial Fraud Detection Systems ​

Author: Muhammad Waleed Gul, Elaheh Homayounvala
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.22607v1 Announce Type: new Abstract: Financial fraud detection relies heavily on centralized machine learning models. This creates serious data privacy risks. Federated Learning (FL) decentralizes data processing, but financial regulations still require models to be transparent. This mean...

📖 Read original article


102. What AstroPT knows about galaxies, and what that can teach us about LLMs ​

Author: UniverseTBD, :, Kshitij Duraphe, Aman Kumar, Michael J. Smith, Shashwat Sourav
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.IM

arXiv:2608.22614v1 Announce Type: new Abstract: Interpretability research increasingly asks when concepts emerge during training and whether linear probes recover real structure, but in language models these claims are hard to validate because language offers little ground-truth ordering of concepts...

📖 Read original article


103. KMGen: A Skill-based Approach for Synthetic Individual Patient Data Generation ​

Author: Jalen Jiang, Chufan Gao, Ethan Rasmussen, Stephen Z. Xie, Jimeng Sun
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22618v1 Announce Type: new Abstract: Individual patient data (IPD) from clinical trials is the substrate for survival modeling, meta-analysis, and safety research, yet IPD is rarely released. Prior work has addressed only half of this gap: reconstructing Kaplan-Meier (KM) curves from publ...

📖 Read original article


104. Learning Generalizable Behaviors for Terminal Agents ​

Author: Yihang Yao, Bo Pang, Xuan Phi Nguyen, Ding Zhao, Shafiq Joty, Semih Yavuz
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22631v1 Announce Type: new Abstract: Terminal agents are a compelling application of large language models (LLMs), with the potential to integrate deeply into users' daily workflows. Reinforcement learning (RL) is a key technique for improving their capabilities, making scalable training ...

📖 Read original article


105. Q-Learning with Stable Infinite-Dimensional Linear Function Approximation ​

Author: Shengbo Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2608.22636v1 Announce Type: new Abstract: Q-learning with linear function approximation can be unstable because an arbitrary approximation architecture need not preserve the Bellman contraction. We develop a stable infinite-dimensional linear function approximation framework for Q-learning fro...

📖 Read original article


106. Mol-JEPA: A multimodal Joint Embedding Predictive Architecture for Molecules ​

Author: Florian Rottach, Sebastian Schieferdecker, William Rudman, Randall Balestriero, Carsten Eickhoff
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22642v1 Announce Type: new Abstract: Despite recent advances in molecular foundation models, several limitations remain, such as chemically invalid augmentations, modality collapse, and incomplete representation of biochemical environments. To address these challenges, we present \textbf{...

📖 Read original article


107. Maximum-distance nonnegative matrix factorization for unmixing highly mixed grain-size distribution data: A generalization of AnalySize ​

Author: Qianqian Qi, Zhongming Chen, Peter G. M. van der Heijden
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22681v1 Announce Type: new Abstract: Nonnegative matrix factorization (NMF) decomposes a nonnegative matrix into the product of two nonnegative matrices. This property makes NMF well suited for unmixing grain-size distribution data, which are inherently nonnegative and have row sums equal...

📖 Read original article


108. Spiking Neural Networks for Continuous Control: Neuromorphic Reinforcement Learning in Conventional Computing ​

Author: Jessica Hunter, Md Maruf Hossain Shuvo, Krishna Roy
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2608.22729v1 Announce Type: new Abstract: Reinforcement learning (RL) algorithms have made strides over the past decade applying them to a wide range of problems and control tasks. However, the deployment of RL on neuromorphic hardware for continuous control tasks remains under-validated. Name...

📖 Read original article


109. MOSH-WM: Mask-Grounded Soft-Hamiltonian Dynamics for Object-Centric World Models ​

Author: Zhekai Wang, Haoxiang Huang, Xiang Liu, Zhikang Chen, Yueqing Sun, Qi Gu, Shiji Zhou, Miao Liu, Sen Cui
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22750v1 Announce Type: new Abstract: Object-centric world models forecast future videos by evolving a set of entity slots, but the variables receiving dynamics supervision are often unconstrained visual features. We introduce \method{}, a mask-grounded soft-Hamiltonian world model that ma...

📖 Read original article


110. LpWM: A Case for Sparse Representations in World Models ​

Author: Yilun Kuang, Yash Dagade, Quentin Le Lidec, Lucas Maes, Randall Balestriero, Yann LeCun
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22764v1 Announce Type: new Abstract: Joint-embedding predictive architectures (JEPAs) learn latent dynamics for planning and avoid representation collapse by matching features to maximum-entropy distributions such as isotropic Gaussians, yielding dense representations. However, it is uncl...

📖 Read original article


111. Learning to Control Coupled-Dynamics Environments with Joint Markov Decision Processes ​

Author: Ege C. Kaya, Aliasghar Pourghani, Mahsa Ghasemi, Vijay Gupta, Abolfazl Hashemi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.22765v1 Announce Type: new Abstract: Coupled-dynamics environments expose the one-step outcomes that would follow from several possible counterfactual actions under a common realization of exogenous randomness. The ordinary Markov decision process formalism allows one to reason about the ...

📖 Read original article


112. Neural Operator based Multi-Field Reconstruction of Inner Solar Boundary State ​

Author: Vignesh Kumar Pandian Sathia, Reza Mansouri, Dustin J. Kempton, Pete Riley, Rafal A. Angryk
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.IM, astro-ph.SR, cs.CV

arXiv:2608.22782v1 Announce Type: new Abstract: The Solar wind is a continuous flow of charged particles emanating from the solar surface and governed by complex, interacting magnetohydrodynamic processes. Accurate specification of inner-boundary conditions is essential for heliospheric modeling and...

📖 Read original article


113. ReCoG: Reciprocal Co-Evolution for Multimodal Graph Learning ​

Author: Rui Xue, Tianfu Wu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22786v1 Announce Type: new Abstract: Multimodal graph learning requires jointly training over graph structure and heterogeneous node attributes, yet existing methods largely decouple these processes: prior multimodal graph neural networks (GNNs) focus on aligning modalities in a shared em...

📖 Read original article


114. Contrastive Representation-Guided Genetic Minority Oversampling for Imbalanced Time-Series Classification ​

Author: Wenbin Pei, Yunrong Hao, Zhen Liu, Guan Wang, Bing Xue, Yiu-Ming Cheung, Qiang Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22804v1 Announce Type: new Abstract: Real-world time-series classification tasks often exhibit class imbalance, which can be extremely severe in some applications. To avoid training biased classifiers on imbalanced data, sampling is one of the most popular data pre-processing techniques b...

📖 Read original article


115. Change Detection in Probability Flow ODE: Online Testing in Diffusion Latent Spaces ​

Author: Artem Kraevskiy, Artem Prokhorov
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.22807v1 Announce Type: new Abstract: A rapidly growing range of sequential data tasks, such as identifying trend reversals in financial markets, auto-segmenting video and audio recordings, detecting changes in movement direction from motion sensors cannot be fully addressed without detect...

📖 Read original article


116. CatchBench: When Can an Agent Failure Be Caught? ​

Author: Yue Zhao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.MA, cs.PF

arXiv:2608.22808v1 Announce Type: new Abstract: When can an agent failure be caught? An audit is usually limited by the record rather than by the method. CatchBench therefore puts one auditor's question to three information states: the declared configuration before a run (PRE), a growing prefix of i...

📖 Read original article


117. SAGE: Stability-Aware Graph-Based Ensemble Feature Selection for Explainable Postpartum Depression Risk Prediction ​

Author: Md. Rokon Islam Emon, Syed Shariar Alam Shuvo, Shahriar Siddique Ayon, Abdullah Al Mamun, Ahnaf Atef Choudhury
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22809v1 Announce Type: new Abstract: Postpartum depression (PPD) poses a major burden on maternal and child health, especially in low- and middle-income countries where prevalence exceeds 19%. Despite advancements in machine learning for PPD prediction, current approaches are limited by o...

📖 Read original article


118. Fairness-Aware Mixture-of-Experts via Subgroup Reweighting and Gate Regularization ​

Author: Sunhee Hwang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22820v1 Announce Type: new Abstract: Deep learning models often produce performance disparities across demographic groups, due to the training data imbalance with respect to sensitive attributes such as gender or age. To address this problem, existing work has explored fair representation...

📖 Read original article


119. DIME: Query-Efficient Framework for Membership Inference on Diffusion Models ​

Author: Tue Do, Daniel Alabi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.22824v1 Announce Type: new Abstract: Membership inference attacks expose whether individual records were used to train a model, yet existing attacks on diffusion models are largely heuristic and can require substantial query budgets. We introduce DIME (Denoiser Ideal Membership Error), a ...

📖 Read original article


120. Hierarchy-Aware Supervised Uncertainty Estimation for Black-box LLM Taxonomic Reasoning ​

Author: Shuting Xie, Nathaniel Lesperance, Graham W. Taylor
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22839v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for scientific decision support, yet reliable confidence estimation remains difficult in black-box settings. We study uncertainty estimation for hierarchical taxonomic reasoning generated by a black-bo...

📖 Read original article


121. RIBOSPAN: A Long-Context RNA Foundation Model for Versatile RNA Modeling ​

Author: Ziyuan Wang, Bohao Tang, Fei Zhang, Shuo Han, Pengfei Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, q-bio.GN

arXiv:2608.22849v1 Announce Type: new Abstract: Full-length RNAs, particularly messenger RNAs, often exceed the context lengths used to pretrain existing RNA foundation models, limiting complete-transcript modeling at single-nucleotide resolution. We present RIBOSPAN, a 1.61-billion-parameter bidire...

📖 Read original article


122. Thinking at the Right Size: Amortized Distillation Across Post-Trained LLMs ​

Author: Yan Zhou, Sara Kangaslahti, Jonathan Geuter, Nihal V. Nayak, Marco Fumero, Francesco Locatello, David Alvarez-Melis
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22854v1 Announce Type: new Abstract: Practical deployment of large language models (LLMs) requires families of post-trained variants---instruction-tuned, reasoning-tuned, and chat-style models---each at multiple sizes to meet diverse latency and memory budgets. Producing each (variant, si...

📖 Read original article


123. Mapping the Concept Landscape: Structural Perception of Global Distributions for Transparent Data Pruning ​

Author: Dongyue Wu, Tao Ma
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.22858v1 Announce Type: new Abstract: Existing data pruning methods predominantly rely on high-dimensional feature embeddings to measure sample importance. However, these compressed vectors often obscure fine-grained semantic interactions, leading to suboptimal coverage of rare semantic co...

📖 Read original article


124. Stochastic Separability of Embedding Manifolds ​

Author: Liqing Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.22874v1 Announce Type: new Abstract: Neurobiological studies and representation learning have observed that representations of objects belonging to the same category in high-dimensional neural spaces exhibit low-dimensional object manifold characteristics, and different object manifolds a...

📖 Read original article


125. The Mask Is Not the Model: Auditing Prefix Invariance in Attention, State-Space, and Hybrid Sequence Models ​

Author: Taebong Kim, Youngsik Hong, Minsik Kim, Sunyoung Choi, Jaewon Jang, Minseo Kim
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22876v1 Announce Type: new Abstract: We formalize prefix invariance: representations at position t must not depend on future inputs. We give a lightweight audit, two forward passes, no training or gradients, that localizes exactly where causality breaks. Attention-mask inspection is incom...

📖 Read original article


126. Safety Hacking in Constrained Best-of-$N$ Inference-time Scaling ​

Author: Akifumi Wachi, Takumi Tanabe, Youhei Akimoto
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CR

arXiv:2608.22915v1 Announce Type: new Abstract: Inference-time pipelines often sample multiple outputs, filter them with a learned safety model, and return the proxy-feasible output with the highest learned reward. We show that this composition creates a two-stage failure: an imperfect safety proxy ...

📖 Read original article


127. A Momentum-Based Variance-Reduced Algorithm for Federated Multiobjective Optimization ​

Author: Yong Zhao, Chunlin You, Minh N. Dao, Zai-Yun Peng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.22945v1 Announce Type: new Abstract: Federated learning has traditionally been formulated as a single-objective optimization problem, primarily focused on maximizing model utility. In real-world applications, however, machine learning models often need to optimize multiple and potentially...

📖 Read original article


128. Stochastic gradient descent with initial regularization ​

Author: Nabil Kahal'e
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2608.22953v1 Announce Type: new Abstract: We analyze a variant of stochastic gradient descent with initial regularization (SGDIR) and derive dimension-free upper bounds on its expected excess risk for the squared loss. In the noiseless case, we obtain new bounds for both averaged and non-avera...

📖 Read original article


129. Do Time-Series Foundation Models Pay Off for Industrial Monitoring? A Cost-Aware Empirical Study ​

Author: Guan-Hua Wen, Kuan-Yu Chen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22968v1 Announce Type: new Abstract: Industrial monitoring models must detect operationally relevant deviations while satisfying target-specific data, calibration, and resource constraints. Time-series foundation models (TSFMs) promise reusable representations and zero-shot forecasts, yet...

📖 Read original article


Author: Filip Kronstr"om, Ross D. King
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22981v1 Announce Type: new Abstract: Knowledge graphs are often accompanied by ontological class hierarchies that encode valuable semantic information, yet many link prediction methods either ignore such hierarchies or incorporate them indirectly through additional graph edges. Recent wor...

📖 Read original article


131. A Physical Response-and-Memory Model for Muon Optimization ​

Author: Yinze Hu, Hongjun Xiang, Xingao Gong, Hongyu Yu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, cond-mat.stat-mech, cs.AI, physics.comp-ph

arXiv:2608.22994v1 Announce Type: new Abstract: Training large language models is costly. How low a loss the same compute can ultimately reach depends on how each step's gradient is converted into a weight update; the rule that performs this conversion is the optimizer. From SGD and AdamW to the rec...

📖 Read original article


132. SplitLite: Low-Rank Residual Compression for Split Learning ​

Author: Tao Li, Yulin Tang, Qi Guo, Xianhao Chen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23018v1 Announce Type: new Abstract: Federated fine-tuning of on-device large language models (LLMs) faces a significant computing burden. To overcome this limitation, split learning (SL) has emerged as a promising solution, which offloads the primary training workload to a powerful serve...

📖 Read original article


133. FedCC: Towards Addressing Label Distribution Skews in Distillation-Based Federated Learning ​

Author: Wenxuan Ye, Onur Ayan, Xueli An, Georg Carle
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23031v1 Announce Type: new Abstract: Federated Learning (FL) enables distributed clients to collaboratively train models without sharing raw data, making it promising for leveraging massive devices in communication networks. In distillation-based FL, each client applies its local model on...

📖 Read original article


134. ST$^2$U: Stateful Test-Time Unlearning via Restricted Knowledge Boundary Control ​

Author: Xunlei Chen, Qinghui Gong, Ruini Xue, Yaodong Hu, Tian Lan, Wenhong Tian
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.23034v1 Announce Type: new Abstract: Controlling restricted knowledge in large language models is essential for model alignment and safe deployment. Test-time unlearning avoids costly retraining and parameter updates by intervening only during inference. However, existing activation-editi...

📖 Read original article


135. Reservoir of Importance: Learning Semi-Structured Sparsity with Differentiable Subset Sampling ​

Author: Ha Dinh, Xuan Duy Ta, Khoat Than, Khac-Hoai Nam Bui
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23048v1 Announce Type: new Abstract: Semi-structured $N$:$M$ sparsity has emerged as a practical direction for accelerating large language models (LLMs). However, existing learnable-mask approaches incur substantial parameter and memory overhead, limiting their scalability to large models...

📖 Read original article


136. Graph Representation Learning of Lightweight IoT Ciphers ​

Author: Jonathan Cook, Sabih ur Rehman, M. Arif Khan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.23054v1 Announce Type: new Abstract: SIMON and SIMECK belong to a family of Lightweight Cryptographic Algorithms (LCAs) based on the Feistel block cipher, designed for Internet of Things (IoT) devices. As with all Feistel ciphers, they are susceptible to differential cryptanalysis, necess...

📖 Read original article


137. Macro-Action Topological Navigation under Noisy Localization using Reinforcement Learning ​

Author: Simon Hakenes, Tobias Glasmachers
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.23055v1 Announce Type: new Abstract: Navigating large, photorealistic 3D apartments from raw pixels is widely considered infeasible for plain reinforcement learning. We build an agent that does it anyway, estimating its own pose from the camera alone. The agent has to reach several target...

📖 Read original article


138. PolyChirp: Multi-Species Birdsong Classification Using TinyML on Low-Power Acoustic Sensors ​

Author: Nathan Duboisset, Zhaolan Huang, Felix Bie{\ss}mann, Roudy Dagher, Antoine Lavandier, Emmanuel Baccelli
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23101v1 Announce Type: new Abstract: Recent progress in the field of TinyML has demonstrated that low-power hardware based on microcontrollers can achieve bird species monitoring in real time based on acoustic sensor data for an entire breeding period on a single battery charge. However, ...

📖 Read original article


139. DeMixPert: Decomposed Response Modeling with Gaussian Mixtures for OOD Single-Cell Perturbation Prediction ​

Author: Jiawen Liu, Xuechenxiao Cao, Yutong Li, Bing Liu, Jiaming Liang, Tinghe Zhang, Xiaoqi Sheng, Hongmin Cai
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23114v1 Announce Type: new Abstract: Predicting transcriptome-wide responses to unseen genetic perturbations remains a major computational challenge because accurate prediction requires recovering both perturbation-specific transcriptional shifts and heterogeneous cellular responses. Exis...

📖 Read original article


140. Activation-Weighted Seeded Residual Coding for Low-Bit LLM Weight Repair ​

Author: Zehao Liu, Chuangchuang Fang, Yang Ren
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.23144v1 Announce Type: new Abstract: Low-bit weight quantization saves storage but leaves errors that degrade language-model quality. We introduce Activation-Weighted Seeded Residual Coding (AWSRC), a compact repair codec for an existing quantization backbone. Given a reconstructed weight...

📖 Read original article


141. Conformal Risk Minimization for Semi-Supervised Domain Adaptation via Optimal Transport ​

Author: Manos Giannopoulos, Yi Shen, Michael M. Zavlanos
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23153v1 Announce Type: new Abstract: In high-stakes healthcare applications, machine learning models are frequently trained on data from one patient population and deployed on another, creating a distribution shift that degrades both accuracy and reliability. Semi-Supervised Domain Adapta...

📖 Read original article


142. When More Modalities Hurt: Modality Dropout for Heavy-Duty Vehicle Engine Diagnostics ​

Author: Adeel Zafar, Slawomir Nowaczyk, Hamid Sarmadi, Saeed Gholami Shahbandi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23161v1 Announce Type: new Abstract: Heavy-duty vehicle diagnostics generate three disconnected data modalities: unstructured multi- lingual service complaints, high-dimensional sensor telemetry with over 80% missing values, and Diagnostic Trouble Codes (DTCs). We investigate whether fusi...

📖 Read original article


143. Counterfactual Transition Graphs: Evaluating Cross-Class Transition Quality ​

Author: Syed Muhammad Hamza Zaidi, Szymon Bobek, Grzegorz J. Nalepa, Myra Spiliopoulou
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23164v1 Announce Type: new Abstract: Counterfactual (CF) explanations for time-series classifiers are usually evaluated one example at a time: what minimal edit flips this single window's prediction? We argue that the more informative question for diagnostic interpretability is structural...

📖 Read original article


144. A Comparative Study of Label-free Representation Quality Metrics in Deep Learning ​

Author: Daniel Richards Arputharaj, Daniel J"onsson, Gabriel Eilertsen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.23182v1 Announce Type: new Abstract: We present a comparative study of label-free metrics for assessing the quality of representations in deep neural networks to understand their reliability under a wide variety of configurations. We group existing label-free metrics into three families b...

📖 Read original article


145. Leveraging Remote Traffic Data for Local Air Pollutant Estimation: A Scenario-Based Machine Learning Study Across London Monitoring Sites ​

Author: Valeria Legaria-Santiago, Amadeo Arguelles, Magdalena Saldana-Perez, Jocelyn Richardson, Marcella Bona
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23219v1 Announce Type: new Abstract: Vehicular traffic is a major source of air pollution; however, the contribution of remotely acquired traffic information to local machine-learning (ML) air-pollution models remains insufficiently characterised. This study evaluates four interpretable t...

📖 Read original article


Author: Peiyang Liu, Xi Wang, Di Liang, Wei Ye
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IR

arXiv:2608.23252v1 Announce Type: new Abstract: As Retrieval-Augmented Generation (RAG) shifts toward diverse portfolio generation, it is stymied by two critical bottlenecks: flawed measurement of evidence utilization, and suboptimal context budget allocation. We resolve both sequentially. To resolv...

📖 Read original article


147. From Multimodal Observation to Interpretable Suggestions: Counterfactual Time-Expanded Relational Modeling of Surgical Teams ​

Author: Vincenzo Marco De Luca, Antonio Longa, Giovanna Varni, Andrea Passerini
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23254v1 Announce Type: new Abstract: In surgery, patient safety is threatened not only by technical issues but also by poor teamwork. However, existing surgical AI-based solutions focus mainly on visual workflow and technical execution, neglecting the modeling of team interactions and mis...

📖 Read original article


148. A Multidimensional Data-Driven Hybrid Transformer Framework for Non-invasive Continuous Blood Pressure Prediction ​

Author: Yuexin Ma, Jingqi Hou, Yuxuan Kang, Zhaoying Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23276v1 Announce Type: new Abstract: Objective. To develop and evaluate a cuffless continuous blood pressure (BP) estimator using temporal physiological and demographic features. We propose a hybrid Transformer framework to estimate diastolic and systolic BP from ECG/PPG-derived feature s...

📖 Read original article


149. How Much Regularization Survives Averaging? Update Masking in Federated Learning ​

Author: Wenhao Yan, Fu Kuroda, Yucheng Jin, Zhenke Chen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23286v1 Announce Type: new Abstract: Federated learning on non-IID data seeks flat minima to generalize across clients, and existing methods borrow sharpness-aware minimization from centralized training. There is a second way to reach flat minima, in which the regularization comes for fre...

📖 Read original article


150. Poisson Subspace Clustering: Focusing on the Essentials in Count Data ​

Author: Collin Leiber, Kai Puolam"aki, Heikki Mannila
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23287v1 Announce Type: new Abstract: Count data represented as a matrix of non-negative integer values, such as contingency tables, are prevalent across diverse domains. When clustering such data sets, specific methods are required, as generic algorithms often fail to consider their uniqu...

📖 Read original article


151. Sigmoid Attention as a Better Substrate for Learned KV Cache Eviction ​

Author: Isaac (Rucheng), Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23296v1 Announce Type: new Abstract: Learned KV-cache eviction often faces a soft-to-hard mismatch: during training, differentiable gates typically attenuate token contributions, whereas inference saves memory only when KV entries are physically removed. We ask whether the attention subst...

📖 Read original article


152. Beyond Point Predictions: Uncertainty-Aware Satellite Poverty Mapping for Public Policy ​

Author: Markus B. Pettersson, James Bailie, Mohammad Kakooei, Eagon Meng, Adel Daoud
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23322v1 Announce Type: new Abstract: Despite their critical importance for policy and research, high-resolution poverty data remain limited across much of Africa. Machine learning (ML) with earth observation (EO) imagery has recently emerged as a way to supplement these data by predicting...

📖 Read original article


153. Towards Actionable Surgical Team Dynamics: from Teamwork to Counterfactual Annotations ​

Author: Vincenzo Marco De Luca, Antonio Longa, Andrea Passerini
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23344v1 Announce Type: new Abstract: Modeling team interactions in high-stakes environments such as operating rooms is critical for understanding how coordination, communication, and individual behaviors shape team performance and safety outcomes. Existing datasets in this domain are ofte...

📖 Read original article


154. Test-Time Adaptation for ECG Classification via SQI-Gated Self-Training and Beat-Rhythm Consistency ​

Author: Wenhan Jiang, Zhipeng Deng, Jiale Zhou, Haolin Wang, Yafei Ou, Yefeng Zheng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23347v1 Announce Type: new Abstract: Deep learning models for electrocardiogram (ECG) classification often suffer from significant performance degradation when deployed in unseen domains due to shifts in acquisition devices and patient populations. Test-time adaptation (TTA) offers a prac...

📖 Read original article


155. Spectrum-Aware Bounds on Invertibility for Privacy-Enhancing Instance Encoding ​

Author: Seokjin Hwang (Ray), Yuting (Ray), Li, Kiwan Maeng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.23382v1 Announce Type: new Abstract: Instance encoding is a popular empirical technique for privacy enhancement when sharing data to an untrusted server. It transforms sensitive data through an encoding process before sharing, with the hope that the encoding process retains utility but ma...

📖 Read original article


156. The Axiomatic Trader: Latent Regularity, Information Budgets, and the Canonical Form of a Quantitative Investment System ​

Author: Jiayu Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, q-fin.PM

arXiv:2608.23416v1 Announce Type: new Abstract: Systematic trading rests on one article of faith: that regularities found in the past persist. We state it as a time-invariant mechanism driven by an unobserved latent state, and show that it leaves a researcher five constants to declare --- the recurr...

📖 Read original article


157. ChebBooster: A Training-Free Approach for Efficient Diffusion Transformer Inference via Chebyshev-Inspired Extrapolation ​

Author: Chengjie Lu, Tianchi Deng, Zhengqi He, Chengwen Luo, Xueliang Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23429v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have shown strong performance in high-fidelity image generation, but their sampling process remains computationally intensive due to full model execution at every timestep. While cache-based acceleration has been explored ...

📖 Read original article


158. Traceable Spectral Inference via Influence Functions: Efficient Data Attribution and Error Proxies for the Ariel Mission ​

Author: Nikki Grens, Lu'is F. Sim~oes, Kai Hou Yip, Theresa Lueftinger
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.IM, stat.ML

arXiv:2608.23458v1 Announce Type: new Abstract: Interpretability is critical for machine learning models deployed in scientific space missions such as ESA's Ariel, where ground truth is unavailable during operations and physical plausibility must be assessed. While most explainable AI methods focus ...

📖 Read original article


159. Diversity-Based Active Learning: An Evaluation of Metric Spaces for Active Learning Selection ​

Author: Siddharth Chilamkur, Dorit S. Hochbaum
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23461v1 Announce Type: new Abstract: With rapid advancement over the last few years, many different methods are now widely used for classification. However, training these models requires substantial labeled data. Active Learning is a potential solution to this problem. Pool-based active ...

📖 Read original article


160. ProxyFormer: A Dual-Stream Proxy Architecture for Ultra-Long Context and High-Resolution Generation ​

Author: Zhongpan Tang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23463v1 Announce Type: new Abstract: The quadratic growth of attention computation and key-value (KV) cache with respect to sequence length is a central bottleneck for ultra-long-context language models and high-resolution generative models. We propose ProxyFormer, a general dual-stream a...

📖 Read original article


161. RAD: Rule-Augmented Relational Anomaly Detection ​

Author: Noah Dahle, Anne Tumlin, Ngoc Tran, Xenofon Koutsoukos, Tyler Derr
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.DB

arXiv:2608.23468v1 Announce Type: new Abstract: Anomaly detection is often applied to data stored in relational databases, yet most existing methods require flattening multiple tables into a single feature matrix. This flattening can obscure entity identity, schema structure, and multi-hop dependenc...

📖 Read original article


162. MetaCaster: Meta-Harness-Optimized Agent for End-to-End Few-Shot Learning of Lightweight Time Series Forecasters ​

Author: ChengAo Shen, Wenchao Yu, Fangyu Wu, Dongjin Song, Hanghang Tong, Dongsheng Luo, Wei Cheng, Haifeng Chen, Jingchao Ni
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23473v1 Announce Type: new Abstract: Time series forecasting (TSF) is evolving toward multimodal and agentic settings, yet using foundation models remains uneconomical in resource-constrained scenarios, where compact, specialized forecasters are more desirable. However, lightweight foreca...

📖 Read original article


163. Provably adaptive sampling with uniform and remasking discrete diffusion models ​

Author: Daniil Dmitriev, Zhihan Huang, Yuting Wei
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, math.ST, stat.ML, stat.TH

arXiv:2608.23554v1 Announce Type: new Abstract: Discrete diffusion models offer a promising alternative to autoregressive generation by enabling parallel updates, but their sampling efficiency can depend strongly on the choice of the forward process and the sampler. For the uniform forward process, ...

📖 Read original article


164. How to Train a Critic Stably and Efficiently ​

Author: Penghui Qi, Xiangxin Zhou, Wee Sun Lee
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.23566v1 Announce Type: new Abstract: Group-based reinforcement learning methods such as GRPO for large language models avoid training a critic by sampling multiple responses for each prompt. A reliable critic could instead estimate token-level advantages from one response, but standard cr...

📖 Read original article


165. A Deep Causal Inference Approach to Measuring the Effects of Forming Group Loans in Online Non-profit Microfinance Platform ​

Author: Thai T. Pham, Yuanyuan Shen
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.IR, cs.LG, q-fin.GN

arXiv:1706.02795v1 Announce Type: cross Abstract: Kiva is an online non-profit crowdsouring microfinance platform that raises funds for the poor in the third world. The borrowers on Kiva are small business owners and individuals in urgent need of money. To raise funds as fast as possible, they have ...

📖 Read original article


166. Reviewing Model Collapse and Countermeasures ​

Author: Xihao Xie, Beichen Hu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.21366v1 Announce Type: cross Abstract: Driven by massive amounts of web-scale data, generative AI (GenAI) has achieved remarkable progress, enabling various applications in diverse sectors. The advances of GenAI have actuated practitioners to use AI-synthesized data for training next-gene...

📖 Read original article


167. PepLLM: ESM-Guided Llama for Structured Protein-Peptide Binding Interface Analysis ​

Author: Hao Qian, Shikui Tu, Lei Xu
Published: 8/25/2026, 4:00:00 AM
Categories: q-bio.BM, cs.AI, cs.LG

arXiv:2608.21367v1 Announce Type: cross Abstract: Protein-peptide interactions are central to cellular regulation and peptide-based drug discovery, yet existing computational methods mainly focus on interaction classification, binding-site prediction, or peptide binder generation. These formulations...

📖 Read original article


168. AI Learning and Conceptual Transfer in the Game of Hidden Rules ​

Author: Christo Mathew, Wentian Wang, Jacob Feldman, Lazaros K. Gallos, Paul B. Kantor, Vladimir Menkov, Hao Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.21372v1 Announce Type: cross Abstract: This report summarizes the work conducted on the Game of Hidden Rules (GOHR), focusing on reinforcement learning agents trained to infer hidden rules from trial-and-error feedback, representation design, rule difficulty analysis, transfer learning, g...

📖 Read original article


169. Agentic Scaffolding Amplifies Sycophantic Behavior in Large Language Models ​

Author: Thantham Jittham
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.MA

arXiv:2608.21377v1 Announce Type: cross Abstract: Sycophancy in large language models, the tendency to prioritize user agreement over truthful responses, has been documented extensively but studied primarily in single-turn settings. This paper investigates a critical question: does subjecting LLMs t...

📖 Read original article


170. RiskWorld: Object-Centric Latent World Modeling for Autonomous Driving Risk Identification ​

Author: Jingzheng Li, Yufei Ge, Qianren Mao, Zhijun Chen, Bing Li, Xingyu Peng, Baochang Zhang, Xianglong Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.21414v1 Announce Type: cross Abstract: Autonomous driving risk identification aims to determine which observed object is likely to become safety-critical to the ego vehicle. Existing approaches typically predict scene-level accidents, infer risk objects indirectly from ego behavior, or ap...

📖 Read original article


171. EditStream: A Unified Autoregressive Framework for Interactive Video Generation and Editing ​

Author: Yuqian Zhou, Zhenghong Zhou, Zongze Wu, Cameron Smith, Richard Zhang, Jiebo Luo, Eli Shechtman, Zhe Lin
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.HC, cs.LG, cs.MM

arXiv:2608.21424v1 Announce Type: cross Abstract: Interactive video generation and editing are becoming increasingly important for creative design. In this report, we introduce EditStream: a unified framework for interactive video generation and editing. EditStream unifies multiple video creation an...

📖 Read original article


172. Constructing Predictive Surgical Path for AI-based Capsulorhexis Skill Transfer ​

Author: Mohammad Javad Ahmadi, Hamid D. Taghirad
Published: 8/25/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2608.21441v1 Announce Type: cross Abstract: Automated training of surgeons is one of the most crucial factors that significantly minimize surgical training risks and expenses. With recent advances in artificial intelligence (AI) knowledge and available data from various surgeries, AI's involve...

📖 Read original article


173. DNA Methylation Profiling in Melanoma: From Lesion Classification to Therapeutic Stratification ​

Author: Jana T. Winterstein, Lukas Heinlein, G"unter Raddatz, Carina Nogueira Garcia, Sarah Haggenm"uller, Christoph Wies, Lucas Schneider, Annemarie Hoffsommer, Tim J. Zeuner, Friedegund Meier, Sarah Hobelsberger, Frank F. Gellrich, Mildred Sergon, Axel Hauschild, Lucie Heinzerling, Justin G. Schlager, Kamran Ghoreschi, Max Schlaak, Franz J. Hilke, Carola Berking, Markus V. Heppt, Michael Erdmann, Sebastian Haferkamp, Konstantin Drexler, Dirk Schadendorf, Wiebke Sondermann, Matthias Goebeler, Bastian Schilling, Daniel B. Lipka, Stefan Fr"ohling, Felix Sahm, Jakob N. Kather, Yuri Tolkach, Jochen S. Utikal, Benjamin Izar, Yevgeniy R. Semenov, Titus J. Brinker
Published: 8/25/2026, 4:00:00 AM
Categories: q-bio.GN, cs.LG

arXiv:2608.21448v1 Announce Type: cross Abstract: DNA methylation provides a stable record of cellular identity, capturing epigenetic programs that distinguish specialized cell states despite a shared genome. Because malignant transformation and tumour progression are accompanied by extensive epigen...

📖 Read original article


174. CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance ​

Author: Erik Thureck, Leo S. R"dian
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.21462v1 Announce Type: cross Abstract: Due to the selection of their training data, large language models (LLMs) perform best on standard-language inputs from languages using the Latin alphabet with large speaker populations, while disadvantaging other language varieties. Nevertheless, th...

📖 Read original article


175. Enhanced Artificial Neural Networks Using QHAdamW in Air Quality Forecasting ​

Author: Mary Joy Daniel Vinas
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.ET, cs.LG, cs.NE

arXiv:2608.21463v1 Announce Type: cross Abstract: The study employed an Artificial Neural Network in combination with the optimized Adaptive Moment Estimation (Adam) algorithm, currently the only AQI forecasting model available in the Philippines. The modified QHAdamW - Quasi-Hyperbolic Momentum (QH...

📖 Read original article


176. Complexity Induction: Compositional Generalization via Structured Label Distortion ​

Author: Aleksandr Abramov
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.21464v1 Announce Type: cross Abstract: We demonstrate that structured distortion of training data - which we term complexity induction - can induce compositional generalization in a standard CNN classifier without architectural modification. Using synthetic images of colored geometric sha...

📖 Read original article


177. Spectral partitioning for $k$-block averaging kernels of finite Markov chains ​

Author: Michael C. H. Choi, Youjia Wang
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT, math.OC, math.PR, stat.CO

arXiv:2608.21466v1 Announce Type: cross Abstract: We develop spectral algorithms for selecting state-space partitions that define averaging kernels for finite, ergodic and reversible Markov chains. For a partition $\mathcal O$, the Gibbs kernel $G_{\mathcal O}$ resamples within the current block fro...

📖 Read original article


178. Gauss--Hermite Quadrature for Gaussian-Mixture Entropy with an Action-Space Hermite Surrogate ​

Author: Jae Wan Shim
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT

arXiv:2608.21467v1 Announce Type: cross Abstract: Gaussian distributions are used to model uncertainty in signals and states, and Gaussian mixtures are often used when the underlying distribution is multimodal. Unlike a single Gaussian, a Gaussian mixture generally has no closed-form expression for ...

📖 Read original article


179. Explainable Adaptive Zero Trust Framework for AWS with Adversarial Robustness Evaluation ​

Author: Om Singh, Yagyaraj Pandey, Nandini Pathak
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CR, cs.CY, cs.LG, cs.NI

arXiv:2608.21477v1 Announce Type: cross Abstract: Cloud environments built on Amazon Web Services face a structural security vulnerability: once a credential passes authentication, the resulting session is often treated as trusted for its entire duration. This assumption fails when credentials are s...

📖 Read original article


180. Magnitude Homology Is the Associated Graded of the Length Filtration ​

Author: Luciano Melodia
Published: 8/25/2026, 4:00:00 AM
Categories: math.AT, cs.CG, cs.LG, cs.LO

arXiv:2608.21479v1 Announce Type: cross Abstract: Magnitude homology is graded by length and knows nothing of persistence. Its persistent refinement knows nothing of where its bars begin and end. We show that the two are one construction: filtering the length nerve by sublevel sets of the length yie...

📖 Read original article


181. A Data-Driven Approach to State Construction in Markov Models ​

Author: Linde Van Gestel, Marie-Anne Guerry, Evy Rombaut
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR, stat.ME

arXiv:2608.21480v1 Announce Type: cross Abstract: A Markov chain is a widely used stochastic process modelling random events over time. These models are built on subsets of the entire dataset, referred to as states, which are considered to be homogeneous regarding transition probabilities. However, ...

📖 Read original article


182. Retrieval Needs Multivectors: An Exponential Separation ​

Author: Mihir Agarwal, Viraj Agrawal, Sabyasachi Basu, Ankit Garg, Kirankumar Shiragur
Published: 8/25/2026, 4:00:00 AM
Categories: cs.IR, cs.DB, cs.LG

arXiv:2608.21494v1 Announce Type: cross Abstract: Recent works have highlighted the expressive limitations of embedding based retrieval models through both theoretical analyses and challenging benchmarks such as LIMIT. While multi-vector embeddings consistently outperform single-vector embeddings, t...

📖 Read original article


183. BeTaL-GBI: Admission-Aware Benchmark Tuning and Full-Stack Verification of Geometric Belief Interfaces ​

Author: Alvin Spivey, Yu Huang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.SE, cs.CR, cs.LG

arXiv:2608.21503v1 Announce Type: cross Abstract: A verification substrate is more credible when exposing errors in its own claims, not just model outputs. GBI-DCSE v3 falsified an architectural claim: the reported Fisher value epsilon ~ 0.066 satisfies the kappa^2 <= 10^4 budget only on the slice [...

📖 Read original article


184. What Neural Network Field Theory Can and Cannot Realise on a Computer ​

Author: Thomas R. Harvey
Published: 8/25/2026, 4:00:00 AM
Categories: hep-th, cs.LG, hep-lat, hep-ph, math-ph, math.MP

arXiv:2608.21523v1 Announce Type: cross Abstract: One aim of neural network field theory is to put a quantum or effective field theory on a computer, with the network ensemble itself as the theory. We ask how far that aim can be pushed for a function class regular enough to be computed with. Our mai...

📖 Read original article


185. Sparse Separable Factor Analysis in the Complex Domain with an Application to Local Field Potential Data ​

Author: Ian Hultman, Kirtikanth Kalapatapu, Yassine Filali, Rainbo Hultman, Sanvesh Srivastava
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO, stat.ME

arXiv:2608.21551v1 Announce Type: cross Abstract: Complex-valued arrays arise in signal processing, where scientific interpretation depends on retaining amplitude and phase information. Existing covariance estimation methods either ignore the multiway organization of such data or rely on real-domain...

📖 Read original article


186. Model-Based Reinforcement Learning for Heterogeneous Multi-Robot Task Assignment Under Distribution Shifts ​

Author: Daniel Garces, Sara Castro, Adrian Haimovich, Byron Crowe, Stephanie Gil
Published: 8/25/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.MA

arXiv:2608.21554v1 Announce Type: cross Abstract: Heterogeneous multi-robot service systems must assign requests to compatible robots, construct feasible schedules, and adapt as new tasks arrive online. Historical data can help anticipate future demand, but relying too heavily on inaccurate predicti...

📖 Read original article


187. Tensor Seeks Layout: Formalizing Layout Selection for ML Compilers ​

Author: Clemens Eisenhofer, Yuwen Jia, Daniel Kroening, Sergey Pupyrev
Published: 8/25/2026, 4:00:00 AM
Categories: cs.PL, cs.DS, cs.LG

arXiv:2608.21555v1 Announce Type: cross Abstract: Modern machine learning compilers select tensor memory layouts to minimize execution cost under hardware constraints. Layout selection is global: an operator may be fastest under one layout while its consumers prefer another, and aligning these prefe...

📖 Read original article


188. A Reproducible, License-Aware Distillation Recipe for CPUDeployable Safety Classification ​

Author: Edson Rodrigues da Cruz Filho, Paulo Ricardo Ferreira Neves, Paulo Henrique Eleuterio Falsetti, Jo~ao Vitor Pavan, Ian Degaspari, Henrique Vieira Laturrague, Patrick Vieira Laturrague, Guilherme Nielsen Dias, Marccello Wilson Perez Berto, Gustavo Voltani Von Atzingen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.21570v1 Announce Type: cross Abstract: Deploying a safety layer for large language models on commodity hardware is constrained by the guards available to do it: current open guard models hold between 1 and 9 billion parameters, are oriented toward the graphics processing unit, and answer ...

📖 Read original article


189. Neural Network Field Theory at Finite Width ​

Author: Christian Ferko, Aaron Mutchler
Published: 8/25/2026, 4:00:00 AM
Categories: hep-th, cond-mat.dis-nn, cs.LG

arXiv:2608.21588v1 Announce Type: cross Abstract: Under mild assumptions, any quantum mechanical (QM) model or quantum field theory (QFT) admits a representation in terms of an ensemble of neural networks with countably many random parameters. We investigate the features of NN-QM and NN-FT models wi...

📖 Read original article


190. Random Hazard Forests ​

Author: Hemant Ishwaran, Eileen M. Hsich, Udaya B. Kogalur, Donald K. K. Lee
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.21597v1 Announce Type: cross Abstract: Clinical data sources such as electronic health records and wearable sensors record patient status repeatedly over follow-up, often at irregular times and on different schedules for different measurements. These data create opportunities for continuo...

📖 Read original article


191. Separating Voice from Age in COPD Screening ​

Author: George P. Kafentzis, Nikoletta Arvaniti
Published: 8/25/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, cs.SD, eess.SP

arXiv:2608.21599v1 Announce Type: cross Abstract: Voice has been proposed as a low-cost screening signal for chronic obstructive pulmonary disease (COPD). COPD is strongly age-associated and voice changes with age, thus such results admit a trivial alternative explanation. We re-evaluate a public su...

📖 Read original article


192. Piecewise Linear Equivariant Maps for Compact Groups ​

Author: Valeriano Aiello
Published: 8/25/2026, 4:00:00 AM
Categories: math.RT, cs.LG

arXiv:2608.21645v1 Announce Type: cross Abstract: Motivated by equivariant neural networks, we study piecewise linear equivariant maps between finite-dimensional real representations of compact groups. We show that all genuinely non-linear piecewise linear behaviour is confined to the subspaces on w...

📖 Read original article


193. MusPyExpress: Extending MusPy with Enhanced Expression Text Support ​

Author: Phillip Long, Hao-Wen Dong, Julian McAuley, Zachary Novack
Published: 8/25/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS

arXiv:2608.21678v1 Announce Type: cross Abstract: Current work in modeling symbolic music primarily relies on representations extracted from MIDI-like data. While such formats allow for modeling symbolic music as sequences of notes, they omit the large space of symbolic annotations common in western...

📖 Read original article


194. Scalable quantum simulation of continuous-time generative models via tensor networks ​

Author: Nathan X. Kodama, L. Andrew Wray, Sam Cochran, Chad Rigetti, Shravan Veerapaneni, Michael J. Keiser
Published: 8/25/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG

arXiv:2608.21700v1 Announce Type: cross Abstract: Continuous-time flow and diffusion models are widely used across many application domains, from large-scale deployment in computer vision and protein folding to emerging adoption for modeling language, time series, and quantum states. After training,...

📖 Read original article


195. The Plan, Not the Decoder: Diagnosing and Repairing Compositional Failure in Reasoning-Augmented Text-to-Image Generation ​

Author: Ashritha Gonuguntla
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2608.21713v1 Announce Type: cross Abstract: Reasoning-augmented text-to-image models such as GoT-R1 emit an explicit textual plan - object names, attributes, and bounding boxes - before generating image tokens. When such a model fails a compositional prompt, is the plan wrong, or is the plan r...

📖 Read original article


196. Guidance for Prior Change via Density Ratio Estimation ​

Author: Yichen Zang, Song Liu, Jiun-Yi Lin
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.21729v1 Announce Type: cross Abstract: Simulation-Based Inference (SBI) serves as a vital framework for parameter inference in scientific fields where simulators involve intractable likelihoods, yet while amortized generative models offer rapid posterior estimation, they are often restric...

📖 Read original article


197. First-Principles Atomistic Structure and Dynamics of Polyethylene During High-Pressure Radical Polymerization via Machine Learning Force Fields ​

Author: Bharatha K. Gunawardana, Teresa Shah, Bicha Azizova, Deepa Ranabhat, Yizhi Song, Akshath Shastri, Srinjoy Ghose, Thomas E. Gartner III, Hsin-Yu Ko
Published: 8/25/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cond-mat.dis-nn, cs.LG, physics.chem-ph

arXiv:2608.21741v1 Announce Type: cross Abstract: Polyethylene (PE) is one of the most commonly used synthetic polymers. While the synthesis and processing protocols for PE are well established, precise experimental assignment of microscopic structures at atomistic resolution (i.e., the position of ...

📖 Read original article


198. HIRA: A Human-in-the-Loop Retrieval-Augmented Cascade for Document Classification in Regulated Industries ​

Author: Shangxuan Tian, Yanhui Chen, Carlos Queiroz
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.IR, cs.LG

arXiv:2608.21792v1 Announce Type: cross Abstract: Document classification in regulated industries is constrained by data residency, limited cold-start labels, scarce review capacity, and costly model-governance procedures. We present HIRA, a training-free, on-premises retrieval-augmented cascade for...

📖 Read original article


199. Hints, Critics, and Teachers: Prior Injection for Sparse-Reward RL in Vision-Language Math Reasoning ​

Author: Qiqian Fu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.21811v1 Announce Type: cross Abstract: Reinforcement learning for vision-language math reasoning starves under sparse reward: on a pool of 20,830 visual-math problems where Qwen2-VL-2B answers 3.6% of rollouts correctly, 85-97% of GRPO rollout groups are entirely wrong and contribute zero...

📖 Read original article


200. PatchGate: Narrowing the Verbalization Gap with Intrinsic Object Inventories in Frozen Vision-Language Models ​

Author: Jihyung Ko, Eunji Jung, Hyeongsub Kim, Ziseok Lee, Jae Won Cho, Sanghyun Jo, Kyungsu Kim
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG

arXiv:2608.21819v1 Announce Type: cross Abstract: Reliable image captioning in Vision-Language Models (VLMs) requires captions to be both precise and complete, avoiding unsupported object mentions while covering visible objects. Existing training-free methods primarily address the former requirement...

📖 Read original article


Author: Jiahao Xie, Guangmo Tong
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG

arXiv:2608.21825v1 Announce Type: cross Abstract: Learning adjacency matrices from node-link images is a fundamental problem for recovering structured graph information from visual observations. Existing methods typically rely on fixed KNN-based heuristics for candidate edge selection and fail to ca...

📖 Read original article


202. The Chase Is the Curriculum, the Capture Anchors the Credit: Pursuit-Evasion Self-Play for Zero-Data LLM Reasoning ​

Author: Jing Yu, Shengchao Chen, Yiyun Tan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.21871v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards has become the dominant recipe for improving large language model reasoning, yet it presumes large human-curated task collections. Zero-data self-play removes this dependency, but existing methods vet le...

📖 Read original article


203. TherMapNet Attention-Guided Runtime Full-Chip Thermal Map Prediction from Performance Metrics ​

Author: Qin Gu, Chaofang Ma, Mingyu Yang, Yipu Zhang, Jiliang Zhang, Wei Zhang, Lin Jiang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AR, cs.LG

arXiv:2608.21887v1 Announce Type: cross Abstract: Runtime thermal management of high-performance chips depends on fast and accurate full-chip thermal maps. Conventional simulators typically estimate power traces from performance metrics first, which adds overhead. This work proposes TherMapNet, an a...

📖 Read original article


204. PhysECD: A Physics-Constrained E(3)-Equivariant Framework for Electronic Circular Dichroism Spectrum Prediction ​

Author: Yi Jiang, Letian Chen, Runhan Shi, Liangzhaoxuan Han, Tong Zhu, Yang Yang
Published: 8/25/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG

arXiv:2608.21892v1 Announce Type: cross Abstract: The electronic circular dichroism (ECD) spectrum is a primary experimental probe for assigning the absolute configuration of chiral molecules, yet interpreting a measured spectrum requires time-dependent density functional theory (TDDFT) calculations...

📖 Read original article


205. EDGE: Experience-Distillation for Guided Exploration in Agentic Reinforcement Learning ​

Author: Can Xie, Yuyi Zhou, Wen Yang, Ziyi zhang, Siyao Song, Yingzhuo Deng, Shuo Ren, Jiajun Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.21946v1 Announce Type: cross Abstract: Reinforcement learning with outcome-based objectives such as GRPO enables LLM-based agents to solve complex, long-horizon tasks, yet the reusable exploration patterns embedded in interaction trajectories are largely discarded after a single policy up...

📖 Read original article


206. ToSCA: Leveraging Hierarchical Reinforcement Learning on Temporal and Strategic Abstractions of Conversational Agents ​

Author: Xiaoyu Wang, Qingqing Gu, Yue Zhao, Teng Chen, Yuqi Cao, Xiaokai Chen, Hongyan Li, Luo Ji
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.HC, cs.LG

arXiv:2608.21969v1 Announce Type: cross Abstract: Humans have multiple levels of temporal abstractions on daily interaction and thinking, such as concept perception and strategic planning. Inspired by this nature, we propose a two-level hierarchical reinforcement learning (RL) framework for conversa...

📖 Read original article


207. Physics-Constrained Neural Flow Maps for Long-Horizon Prediction of Spin Dynamics ​

Author: Haoen Feng, Shenglan Yuan, Shirong Lin
Published: 8/25/2026, 4:00:00 AM
Categories: cond-mat.mes-hall, cs.LG

arXiv:2608.22006v1 Announce Type: cross Abstract: Conventional simulation of current-driven magnetization relies on fine-step integration of the spin-transfer-torque Landau--Lifshitz--Gilbert equation, creating a computational bottleneck in parameter sweeps and control searches. In this work, we pro...

📖 Read original article


208. Barycentric Fused Gromov-Wasserstein Balancing for Causal Inference under Multiple Treatments ​

Author: Yuki Murakami, Takumi Hattori, Kohsuke Kubota
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ME, cs.AI, cs.LG, stat.ML

arXiv:2608.22024v1 Announce Type: cross Abstract: Estimating heterogeneous single and interaction treatment effects from observational data under multiple simultaneous treatments is crucial for decision-making. To mitigate estimation variance, previous studies balance representation distributions be...

📖 Read original article


209. One-Step Evolution for Long-Time Extrapolation: An Error-Bound-Informed and Prior-Guided Neural Residual Framework for Autonomous PDEs ​

Author: Maqun Zhang, Feng Gao, Wankun Chen, Hui Yu, Yanhai Gan, Junyu Dong
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.22026v1 Announce Type: cross Abstract: Accurate simulation of the long-time evolution of systems governed by partial differential equations (PDEs) is central to scientific computing. Among existing deep learning?based approaches for solving PDEs, neural operators typically rely on extensi...

📖 Read original article


210. Structured Learning on Mapper Representations ​

Author: George Babus, Farzana Nasrin
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.22044v1 Announce Type: cross Abstract: Modern machine learning (ML) methods are highly effective for prediction tasks, but many commonly used representations reduce complex data to fixed dimensional embeddings that may suppress multiscale structural organization. The Mapper algorithm from...

📖 Read original article


211. Discovering Dual-Origin Slow Wind from Solar Orbiter with Self-Supervised Contrastive Learning ​

Author: Henry Han, Jorge Yero Salazar
Published: 8/25/2026, 4:00:00 AM
Categories: astro-ph.SR, cs.AI, cs.LG

arXiv:2608.22065v1 Announce Type: cross Abstract: Whether the slow solar wind originates from one coronal source or two distinct channels remains a central open question in heliophysics. Resolving this requires unsupervised separation of two populations that arrive at nearly the same bulk speed and ...

📖 Read original article


212. ADMIL: Attention-Distilled Multiple Instance Learning for Selective Foundation Model Inference in Pathology ​

Author: Duncan Stothers, Ren-Chin Wu, William Lotter
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.22066v1 Announce Type: cross Abstract: Attention-based multiple instance learning (ABMIL) using pathology foundation model embeddings is effective for slide-level tasks, but exhaustive inference requires applying a large image encoder to every foreground tile despite the subsequent attent...

📖 Read original article


213. Inferring Action from Future Latent State for Robotic Manipulation ​

Author: Fenghao Lei, Zhixiong Huang, Long Yang, Jiabao Chen, Jie Cheng, Peilin Huang, Han Fu, Zhuo Li, Xiaoxue Ren
Published: 8/25/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG

arXiv:2608.22067v1 Announce Type: cross Abstract: World-Action Models (WAMs) build robot control on video-generation backbones, which jointly predict dense future visual trajectories and robot actions. We argue that video generation is an unnecessary intermediate objective for world-action modeling....

📖 Read original article


214. Cross-Temperature Defect Identification in Atomistic Simulations via Multi-Level Domain Alignment ​

Author: Yating Fang, Jungmin Kim, Qian Qian Zhao, Pallavi Biswas, Joshua M. Gonjon, Ryan B. Sills, Ahmed Aziz Ezzat
Published: 8/25/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG, physics.comp-ph

arXiv:2608.22074v1 Announce Type: cross Abstract: Identifying atomic defects at elevated temperature is difficult because thermal fluctuations blur the local symmetry that both geometric heuristics and supervised classifiers rely on: trustworthy labels exist in low-temperature reference configuratio...

📖 Read original article


215. Autonomous Cyber Defense: Real-Time Attack Detection and Mitigation in Software-Defined Networks Using Machine Learning ​

Author: Alexandre Amaral, Fernando Moro, Ana Malheiro
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.22075v1 Announce Type: cross Abstract: Adversaries now move faster than manual response processes can absorb. The average eCrime breakout time, that is, the interval between initial access and the first lateral movement to another host, fell to 29 minutes in 2025, a 65% increase in speed...

📖 Read original article


216. When More References Hurt: Contamination-Aware DINOv2 Memory Banks for Few-Shot Steel Defect Detection ​

Author: Hannaneh Kalantari, Javad Khoramdel
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.22082v1 Announce Type: cross Abstract: Patch-memory anomaly detectors assume that their reference bank is normal, an assumption that is difficult to guarantee when additional industrial images are unverified. We study whether a few trusted normal images can safely recover useful normal pa...

📖 Read original article


217. Semantic Reasoning Denoising: Correcting Language Model Reasoning with Semantic Operators ​

Author: Yujiao Yang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.22090v1 Announce Type: cross Abstract: Large language models can produce fluent reasoning traces whose local semantic errors propagate to an incorrect conclusion, while unconstrained self-correction may preserve, amplify, or introduce errors. Existing diffusion language models provide ite...

📖 Read original article


218. Pretreatment DCE-MRI Resolves Response Quality Within Pathologic Endpoints in Neoadjuvant Breast Cancer ​

Author: Dattatreya Kantha, Murray H. Loew
Published: 8/25/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, q-bio.QM

arXiv:2608.22097v1 Announce Type: cross Abstract: Pathologic complete response (pCR) is a strong neoadjuvant endpoint, yet 5-15% of complete responders recur and clinical/genomic variables do not reliably identify them. We tested whether pretreatment dynamic contrast-enhanced MRI entropy - intratumo...

📖 Read original article


219. Development and Feasibility Evaluation of an Edge AI as Medical Device System for Breast Cancer Multidisciplinary Team Meetings ​

Author: Aarzoo Dhiman, Farzana Haque, Kartikae Grover, Lydia Brian Smith, William Stephen Jones
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.22108v1 Announce Type: cross Abstract: Breast Cancer Multidisciplinary Team (MDT) meetings manage increasingly complex cases under considerable time pressure, and documentation requirements can reduce clinical efficiency and decision quality. Existing AI based MDT workflows rely on cloud-...

📖 Read original article


220. Lexical Perturbations Disrupt LLM Reasoning: An Empirical Study of Attention Diversion ​

Author: Jiaqian Zhu, Yang Zhang, Junhua Ding, Xiaowei Yu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.22140v1 Announce Type: cross Abstract: Large Language Models (LLMs) achieve strong reasoning performance, but their robustness to realistic lexical corruption remains poorly understood. We evaluate four open-weight instruction-tuned models and frontier models across four reasoning benchma...

📖 Read original article


221. MCP-Universe RL: A Framework for Training MCP Tool-Use Agents via Reinforcement Learning ​

Author: Ziyang Luo, Yan Yang, Xiangru Jian, Ziji Shi, Xiaoqiang Lin, Jun Hao Liew, Silvio Savarese, Junnan Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.22167v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become an effective way to improve the tool-use ability of large language models (LLMs), but most existing RL frameworks stop at the policy update. For every new domain, the user is left with two hard systems problems:...

📖 Read original article


222. Role-Specialized Mixture-of-Agents with Open-Weight LLMs for Clinical Prediction ​

Author: Jun Hou, Yi Fang, Xuan Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.22176v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly applied to clinical prediction tasks such as in-hospital mortality and readmission from electronic health records (EHRs). Privacy and compliance constraints motivate systems that can be deployed locally, ...

📖 Read original article


223. Token-Level Likelihood-Array Regression for Membership Inference and AI-Generated Text Detection ​

Author: Jiajun Sun, Zhanrui Cai
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.22179v1 Announce Type: cross Abstract: Membership inference asks whether a text was used to train a language model, whereas AI-generated text detection asks whether it was generated by a language model rather than written by a human. Existing likelihood-based methods typically compress to...

📖 Read original article


224. VERDICT: Agreement Beats Pixel-Space Verification in Real-Document OCSR ​

Author: Yani Guan, Dengpan Dong, Shuang Luo, Zi Wei, Joah Han, Dan Hannah, Yumin Zhang, Qichao Hu, Kang Xu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.IR, cs.LG

arXiv:2608.22183v1 Announce Type: cross Abstract: Optical Chemical Structure Recognition (OCSR) converts 2D molecular depictions in the published literature into SMILES, and is increasingly important for constructing large-scale chemical training datasets. Automation at that scale requires identifyi...

📖 Read original article


225. Efficient Regression Models for Scan Statistics ​

Author: Gazi Abdur Rakib, Tristan Ashton, Ryan A. Loomis, Brian S. Mason, Eric J. Murphy, Ci Xue, Jeff M. Phillips
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ME, cs.LG

arXiv:2608.22201v1 Announce Type: cross Abstract: We introduce a new class of regression models for scan statistics on real-valued signals. These allow for improved fitting of non-stationary signals to contrast with the interval anomalies identified by the scan statistics. Our models can represent g...

📖 Read original article


226. Spending Scarce Confirmatory PET Measurements: Target-Aligned Validation in A4/LEARN ​

Author: Eliuvish Han Cui
Published: 8/25/2026, 4:00:00 AM
Categories: stat.AP, cs.AI, cs.LG, stat.ML

arXiv:2608.22223v1 Announce Type: cross Abstract: Anti-amyloid therapies and blood-based biomarkers are changing Alzheimer disease workups into a two-stage measurement workflow: screen broadly with cheaper information, then spend scarce confirmatory amyloid measurements where they support the decisi...

📖 Read original article


227. MRMAD: A Multi-Round Multi-Audio Benchmark for Evaluating Acoustic Degradation Perception in Large Audio-Language Models ​

Author: Yize Li, Ningyuan Yang, Sile Yin, Sindhuja Thogarrati, Sung-En Chang, Andrew C. Singer, Xue Lin, Chuan-Che Huang, Shuo Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS

arXiv:2608.22236v1 Announce Type: cross Abstract: Large audio-language models (LALMs) have shown promising progress in understanding speech, music, and general sound events, yet their ability to reason about how audio signals are degraded remains underexplored. Existing benchmarks primarily evaluate...

📖 Read original article


228. Improving Few-Step Language Flows with Untied Self-Conditioning ​

Author: Bocheng Li, Linli Xu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.22244v1 Announce Type: cross Abstract: Flow-matching language models refine all token positions in parallel and can trade sampling steps for latency, yet generation quality still degrades sharply with few sampling steps. We trace a source of this degradation to a train--inference mismatch...

📖 Read original article


229. Sharp Barron Regularity Results for Coulombic Many-Electron Wave Functions ​

Author: Pingbing Ming, Hao Yu
Published: 8/25/2026, 4:00:00 AM
Categories: math.AP, cs.LG, cs.NA, math.NA

arXiv:2608.22252v1 Announce Type: cross Abstract: We establish sharp Barron regularity for Coulombic many-electron wave functions after extraction of the universal cut-off Jastrow factors. Following the factorization of Fournais et al.~\cite[Definition~1.4]{FournaisEtAl2005}, for a Coulombic eigenfu...

📖 Read original article


230. GAN-Diff : Coupling Pretrained WGAN-GP Features with Conditional Diffusion U-Nets ​

Author: Saif Ahmed, Ashadulla Hil Galib, S. M. Riaz Rahman Antu, Ahmed Faizul Haque Dhrubo, Souvik Pramanik, Mohammad Abdul Qayum, Mohsin Sajjad, Mohammad Ashrafuzzaman Khan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.22272v1 Announce Type: cross Abstract: Generative adversarial networks (GANs) can provide efficient image generation, while diffusion models offer high-quality image restoration but require iterative sampling. This paper presents a hybrid GAN-guided diffusion framework that uses a pretrai...

📖 Read original article


231. Length-Adaptive Decoding for Masked Diffusion Machine Translation ​

Author: Yan Zhan, Mengkai Hou, Wanting Zhang, Zhijun Gao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.22274v1 Announce Type: cross Abstract: Machine translation tests masked diffusion language models (dLLMs) because every source token must be rendered faithfully, while fixed canvas decoding must choose target length before denoising. Existing masked diffusion decoding work mainly studies ...

📖 Read original article


232. Learning from the Test: Self-Referential Differential Testing for Deep RL Agents ​

Author: Junda He, Jieke Shi, Zhou Yang, Mingfei Cheng, David Lo
Published: 8/25/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2608.22284v1 Announce Type: cross Abstract: Deep Reinforcement Learning (DRL) has achieved significant success in complex decision-making problems. As DRL systems are increasingly deployed in real-world applications, ensuring their quality and reliability is paramount. Current works primarily ...

📖 Read original article


233. The spatial anatomy of urban wildfire vulnerability: a spatially validated GeoAI framework reveals the roles of building density and vegetation moisture in structure loss during the 2025 Palisades Fire ​

Author: Parastoo Farajpoor, Mohammadreza Narimani
Published: 8/25/2026, 4:00:00 AM
Categories: physics.geo-ph, cs.LG, eess.IV

arXiv:2608.22293v1 Announce Type: cross Abstract: Urban wildfire resilience depends on interactions among built form, vegetation condition, and extreme fire weather, yet city-scale risk models often overlook whether predictive skill transfers across neighborhoods. We developed a spatially validated ...

📖 Read original article


234. LLM Evaluation on Unseen Questions: Contextual Multidimensional IRT Model ​

Author: Ergan Shang, Weijing Tang, Yinqiu He
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.22295v1 Announce Type: cross Abstract: Evaluation of large language models (LLMs) increasingly requires predicting how a model will perform on new questions or tasks before collecting large amounts of new annotations. This problem is challenging because question difficulty, scenario, and ...

📖 Read original article


235. Recovering Weighted Tangent Geometry from a Single-Scale Score Field ​

Author: Ziqi Zhao, Qingjian Ni
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.22334v1 Announce Type: cross Abstract: Near a smooth data manifold, one tangent space summarizes local geometry. At a branch point, the corresponding first-order object is instead a measure over tangent directions, whose normalized masses record the local share of each branch under the ch...

📖 Read original article


236. Where Cognition Lives: Dissecting Emergent from Computed Function in a Minimal Complete Cognitive Architecture ​

Author: Francisco M. Arrabal-Campos, Francisco G. Montoya, Alfredo Alcayde, Ignacio Fern'andez
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.22347v1 Announce Type: cross Abstract: A cognitive architecture is more than the module that reasons: it must also decide how long to think and what deserves the effort. We built a minimal but complete system - a recurrent reasoner with adaptive halting, a homeostatic control field, and a...

📖 Read original article


237. Dataset Complexity Shapes Finite-Distance Loss Geometry in Neural Networks ​

Author: Jaeyong Bae, Hawoong Jeong
Published: 8/25/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cond-mat.stat-mech, cs.LG

arXiv:2608.22361v1 Announce Type: cross Abstract: Finite datasets can share the same size and low-order statistics while differing strongly in structural complexity. We connect this dataset complexity to loss-landscape geometry by pairing local label mixing across neighborhood scales with local entr...

📖 Read original article


238. Precision-Aware Variable Bit Processing Elements for Hardware-Efficient Systolic Array Designs ​

Author: Dantu Nandini Devi, Madhav Rao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AR, cs.ET, cs.LG, eess.IV

arXiv:2608.22378v1 Announce Type: cross Abstract: Systolic arrays (SAs) have emerged as prominent hardware accelerators for matrix operations in deep learning, while floating point number formats enable precision control across computational domains. This research investigates approximate computing ...

📖 Read original article


239. ProBel: Propaganda Detection with Techniques, Spans, and Explanations ​

Author: Mohamed Bayan Kmainasi, Ali Ezzat Shahroor, Elisa Sartori, Giovanni Da San Martino, Firoj Alam
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.22388v1 Announce Type: cross Abstract: Propaganda detection includes several related prediction levels, ranging from sentence-level decisions to technique classification and span identification. However, it remains unclear how supervision at these levels interacts when learned jointly acr...

📖 Read original article


240. KONTOGRAPH: Verified Point-in-Time Feature Consistency and Amortised Explanation for Real-Time Anti-Money Laundering under a 200 ms Decision Budget ​

Author: Ahmed Abolfadl
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG, cs.SE

arXiv:2608.22389v1 Announce Type: cross Abstract: Regulation (EU) 2024/886 obliges European payment service providers to settle euro credit transfers in under ten seconds, around the clock. This removes both the overnight batch window in which anti-money-laundering (AML) analytics traditionally ran ...

📖 Read original article


241. Arbitrage-Aware Multi-Step Forecasting of Implied Volatility Surfaces: Modelling Surface Trajectories Using Latent Diffusion ​

Author: Dominik Manuel Buchegger, Lukas Gonon
Published: 8/25/2026, 4:00:00 AM
Categories: q-fin.MF, cs.LG

arXiv:2608.22478v1 Announce Type: cross Abstract: Implied volatility surfaces summarise the option market and are central to many financial applications. Forecasting their future evolution requires modelling two-dimensional geometry, temporal dependence, and predictive uncertainty while preserving e...

📖 Read original article


242. GTA-RAG: Graph-Trajectory-Augmented Reinforcement Learning for Multi-Turn Retrieval-Augmented Reasoning ​

Author: Jun Chen, Yongchao Liu, Pengyu Qiu, Jiajun Zheng, Juelu Zhang, Yujie Zeng, Qin Zhang, Ziyue Qiao, Xiao Luo
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.22479v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) enables LLMs to access external knowledge for answering knowledge-intensive questions. For complex multi-hop questions, multi-turn retrieval-augmented reasoning extends RAG into an iterative process that repeatedl...

📖 Read original article


243. Interpretable statistical feature engineering for early disruption prediction in the short pulse ADITYA tokamak ​

Author: Jyoti Agarwal, Kavit Patel, Bhaskar Chaudhury, Abhishek Sharma, Shrichand Jakhar, Manika Sharma
Published: 8/25/2026, 4:00:00 AM
Categories: physics.plasm-ph, cs.LG, physics.data-an

arXiv:2608.22515v1 Announce Type: cross Abstract: Reliable early disruption prediction is critical for the safe operation and real-time control of tokamaks. However, machine learning based prediction frameworks have predominantly targeted medium and long pulse devices, with comparatively limited att...

📖 Read original article


244. Scaling Curriculum Learning For Autonomous Driving ​

Author: Cevahir Koprulu, David Paz, Feng Tao, Yuliang Guo, Xinyu Huang, Ufuk Topcu, Liu Ren
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.22549v1 Announce Type: cross Abstract: Batched simulators for autonomous driving have recently enabled training reinforcement learning (RL) agents at scale, encompassing thousands of traffic scenarios and billions of interactions within a matter of days. Although such high-throughput feed...

📖 Read original article


245. Model-Consistent Byzantine-Resilient Decentralized Federated Learning for Collaborative Missions ​

Author: Yue Li, Sudip Bhujel, Cameron Lira, Ning Wang, Yang Xiao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.DC, cs.CR, cs.LG

arXiv:2608.22552v1 Announce Type: cross Abstract: Decentralized federated learning (DFL) is a promising paradigm for autonomous nodes to collaboratively train AI models without relying on a central server. However, existing DFL solutions do not guarantee global model consistency, a critical requirem...

📖 Read original article


246. Neighbor-embedded Graph Neural Network-based Crowd Delivery Traffic Management in Smart City ​

Author: Kishu Gupta, Deepika Saxena, Ashutosh Kumar Singh, Chung-Nan Lee
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.22555v1 Announce Type: cross Abstract: The significant upsurge in vehicle traffic presents a considerable challenge in the pursuit of smart mobilization and transportation (SMT) worldwide. Current approaches primarily focus on vehicular traffic management through congestion prediction but...

📖 Read original article


247. Two-level domain-decomposition AdaGrad method for scalable training of graph neural networks ​

Author: Laurynas Varnas, Julien Herrmann, Alexander Heinlein, Serge Gratton, Alena Kopani\v{c}'akov'a
Published: 8/25/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA

arXiv:2608.22575v1 Announce Type: cross Abstract: Graph neural networks (GNNs) have emerged as a powerful framework for learning from graph-structured data. However, their efficient training remains challenging, particularly in distributed computing environments. This challenge arises from the use o...

📖 Read original article


248. Sparse Additive Off-Policy Evaluation for Reinforcement Learning with Potentially Limited Number of Trajectories ​

Author: Tuoyi Zhao, Chengchun Shi, Zhengling Qi, Lan Wang
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2608.22595v1 Announce Type: cross Abstract: We develop a new framework for flexible, nonlinear, and interpretable off-policy evaluation for infinite-horizon reinforcement learning. To handle large state spaces and support transparent decision-making, we model the Q-function using a nonlinear f...

📖 Read original article


249. Scale-invariant Optimal Sampling for Rare-events Data with Sparse Models ​

Author: Jing Wang, HaiYing Wang, Qiang Zhang, Hao Helen Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.22597v1 Announce Type: cross Abstract: Subsampling is effective in tackling computational challenges for massive data with rare events. Overly aggressive subsampling may adversely affect estimation efficiency, and optimal subsampling is essential to mitigate the information loss. However,...

📖 Read original article


250. GET: Generative Embedding Translation for Medical Image Segmentation ​

Author: Md Maklachur Rahman, Md Hasan Al Banna, Saraf Anjum, Mahmudul Hasan, Tracy Hammond
Published: 8/25/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2608.22619v1 Announce Type: cross Abstract: Generative segmentation provides an alternative to direct pixel-wise prediction by operating on learned latent representations, but effective image-to-mask translation must preserve target structure while remaining computationally efficient. We propo...

📖 Read original article


251. NeuroPrefetcher: Storage-Aware Sparse LLM Inference via Delta Prefetching ​

Author: Nobel Dhar, Md Romyull Islam, Xuechen Zhang, Gongjin Sun, Sahidul Islam, Bobin Deng, Kun Suo
Published: 8/25/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.22643v1 Announce Type: cross Abstract: Deploying large language models on edge devices is increasingly limited by a widening gap between model size and available memory. Existing approaches such as quantization, smaller models, and offloading can raise the effective memory limit, but they...

📖 Read original article


252. Advanced LLM-Enhanced Intent-Based 5G Network Management using Dynamic Semantic Routes ​

Author: Thomas Benton Townsend, Dimitrios Michael Manias
Published: 8/25/2026, 4:00:00 AM
Categories: cs.NI, cs.LG, cs.SY, eess.SY

arXiv:2608.22644v1 Announce Type: cross Abstract: As the use of Artificial Intelligence (AI) and Large Language Models (LLMs) is becoming common in everyday applications, their ability to interpret natural language has increased significantly. An emerging application of AI is integration with networ...

📖 Read original article


253. Lightweight Multi-scale Hierarchical Anomaly Detection and Localization for Geospatial Big Data Applications at the Edge ​

Author: Thomas Benton Townsend, Joshua Bean, Benjamin K Tkach, Narcisa Gabriela Pricope, Dimitrios Michael Manias
Published: 8/25/2026, 4:00:00 AM
Categories: eess.SP, cs.LG, cs.SY, eess.SY

arXiv:2608.22648v1 Announce Type: cross Abstract: As an increasing number of critical applications, including environmental, emergency, meteorological, and agricultural, rely on real-time anomaly detection in geospatial data streams, challenges related to the storage, processing, and communication o...

📖 Read original article


254. EGAMA-RC: Risk-Calibrated Evidence-Gated Adaptive Malware Analysis for Robust and Interpretable Memory-Forensic Triage ​

Author: Isaac Kofi Nti
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.22721v1 Announce Type: cross Abstract: Machine-learning malware detectors often achieve high clean-data accuracy, but operational triage also requires evidence about uncertainty, novelty, robustness, interpretability, latency, and review cost. This paper presents EGAMA-RC, a risk-calibrat...

📖 Read original article


Author: Jiacheng Ding, Xiaofei Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.DB, cs.LG

arXiv:2608.22727v1 Announce Type: cross Abstract: State-of-the-art temporal-link-prediction (TLP) models are, in essence, multi-channel information aggregators: they combine an interaction-history channel, a time-encoding channel, and a structure channel. The first two have been refined relentlessly...

📖 Read original article


256. Generative Neural Networks for Sinkhorn Distributionally Robust Hypothesis Testing ​

Author: Fenglin Zhang, Teyan Liu, Jie Wang
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC

arXiv:2608.22746v1 Announce Type: cross Abstract: This paper studies the Sinkhorn distributionally robust hypothesis testing (SDRHT) problem, seeking a robust detector against least-favorable distributions in Sinkhorn discrepancy-based ambiguity sets centered at the empirical distributions. Existing...

📖 Read original article


257. XTC: Head-Aware Sampling by Excluding Top Choices ​

Author: Philipp Emanuel Weidmann, Allen Roush, Judah Goldfeder, Sanjay Basu, Ravid Shwartz-Ziv
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.22758v1 Announce Type: cross Abstract: Standard decoding rules for autoregressive language models promote diversity by rescaling the full next-token distribution or truncating its low-probability tail. These strategies overlook a common regime of open-ended generation in which several con...

📖 Read original article


258. Don't Repeat Yourself: Stopping Verbatim Loops at Sampling Time ​

Author: Philipp Emanuel Weidmann, Allen Roush, Judah Goldfeder, Sanjay Basu, Ravid Shwartz-Ziv
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.22761v1 Announce Type: cross Abstract: Large Language Models generate text autoregressively, but open-ended generation is prone to verbatim looping, in which models repeat spans already present in context. Standard defenses such as repetition, presence, and frequency penalties and n-gram ...

📖 Read original article


259. TailSieve: Partial-Rollout-Guided Tail Routing for LLM Rollouts ​

Author: Tianqi Xu, Lu Lv, Haoyang Huang, Wenjie Huang, Zhanming Shen, Yuhao Shen, Baolin Zhang, Xinyi Hu, Shuang Ge, Jun Dai, Tianyu Liu, Suorong Yang, Zhikai Li, Ye Bai, Jun Zhang, Lei Chen, Yue Li, Mingchen Wan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.22788v1 Announce Type: cross Abstract: Large-scale rollouts have become a core component of modern LLM systems, spanning reinforcement learning (RL) post-training, on-policy distillation (OPD), and sampling-heavy evaluation pipelines. Unlike online serving, which is typically optimized fo...

📖 Read original article


260. GuidedFlow: An Attention-Guided Framework for Anomaly Detection in Additive Manufacturing ​

Author: Sosmita Paul, Krishna Roy
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.22789v1 Announce Type: cross Abstract: Additive Manufacturing (AM) plays a vital role in the ongoing industrial revolution. However, quality control remains crucial and challenging due to printing defects or potential cyber-physical intrusions. Image or video-based anomaly detection is a ...

📖 Read original article


261. Beyond the Harness: End-to-End Optimization of Context Artifacts for Enterprise Text-to-SQL ​

Author: Kate Gwimm, Carson Eisenach
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.22830v1 Announce Type: cross Abstract: Deploying LLMs for enterprise Text-to-SQL is bottlenecked less by the model than by what context reaches it: business logic spans thousands of tables, and no model can ingest a full catalog at once. We argue that the most effective place to intervene...

📖 Read original article


262. Mirror descent algorithms with logarithmic barriers ​

Author: Alberto De Marchi, Yura Malitsky, Adrien B. Taylor
Published: 8/25/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.NA, math.NA

arXiv:2608.22834v1 Announce Type: cross Abstract: This work derives convergence guarantees for mirror descent and proximal mirror descent algorithms when a logarithmic barrier is used as a distance-generating function. Standard approaches cannot be applied when the solution lies on the boundary, whe...

📖 Read original article


263. A Commutator Framework for Selective Spectral Alignment in Deep Neural Networks ​

Author: Kaj Nystr"om
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.22910v1 Announce Type: cross Abstract: We develop a finite-width geometric framework describing how learned feature geometries are organized, transported, and selectively aligned in deep neural networks. Incompatibility among weight-generated covariance, gates, and backward sensitivities ...

📖 Read original article


264. Beyond Observed Auxiliary Relations: Environment-Conditioned Modeling for Multi-Behavior Recommendation ​

Author: Seunghan Lee, Hyunsik Yoo, Jian Kang, Susik Yoon, SeongKu Kang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.22920v1 Announce Type: cross Abstract: Multi-behavior recommendation (MBR) leverages auxiliary behavioral signals, such as clicks and add-to-cart, to enhance target behavior prediction like purchases. While recent graph neural network-based approaches have achieved strong performance by s...

📖 Read original article


265. Channel-Token Attention for Reliable Dynamic Spectrum Access under Bursty Primary-User Traffic ​

Author: Krishna Acharya, Dinanath Padhya, Utsab Dahal, Ashish Kandel, Binod Sapkota
Published: 8/25/2026, 4:00:00 AM
Categories: cs.NI, cs.LG, eess.SP

arXiv:2608.22992v1 Announce Type: cross Abstract: Dynamic spectrum access must coordinate secondary users under bursty primary-user activity while preserving packet reliability and delay. We present TACAN, a centralized policy that represents each channel as a token containing occupancy history and ...

📖 Read original article


266. Neural Boltzmann Equations ​

Author: Jonas Spinner, Jack Shergold
Published: 8/25/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, stat.ML

arXiv:2608.23022v1 Announce Type: cross Abstract: The dynamics of particles in the early universe are described by Boltzmann equations, which involve high-dimensional phase-space integrals. Classical approaches use quadrature integration and evolve the system on a fixed momentum grid, which scales p...

📖 Read original article


267. AutoSaddler: Automatic Harness Optimization with Durable Updates from Agent Execution Traces ​

Author: Sungho Park, Wonjoong Kim, Rongyuan Tan, Jue Zhang, Wook-Shin Han, Pengfei Gao, Chanyoung Park, Yongqiang Yao, Rao Fu, Elsie Nallipogu, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA, cs.SE

arXiv:2608.23041v1 Announce Type: cross Abstract: LLM agents remain unreliable on long-horizon tasks, where small local failures can compound over extended interactions and lead to overall task failure. Although external harnesses can substantially improve robustness, harness design remains a manual...

📖 Read original article


268. When a neural surrogate cannot accelerate a solver: runtime share, closed-loop drift, and the economics of uncertainty gating in a stiff coupled simulation ​

Author: L. Th"ummler, T. Kuroda
Published: 8/25/2026, 4:00:00 AM
Categories: astro-ph.IM, astro-ph.HE, cs.LG, physics.comp-ph

arXiv:2608.23075v1 Announce Type: cross Abstract: Learned surrogates for expensive inner solver blocks are a widely pursued route to faster multiphysics simulation. We report a controlled, end-to-end negative result and identify three structural barriers, none of them a deficiency of the network we ...

📖 Read original article


269. Partial-Moment PINNs for Caldeira--Leggett Parameter Learning in Quantum Brownian Motion ​

Author: Krishna Bhatia
Published: 8/25/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.23093v1 Announce Type: cross Abstract: We study parameter recovery in the Caldeira--Leggett (quantum Brownian) oscillator from partial moment traces. Our model is a moment-level PINN that predicts the five first/second moments and enforces the linear CL/HPZ ODEs by automatic differentiati...

📖 Read original article


270. One Inverse Step is a Convex Program: Bayes-Limit Calibration of Diffusion Inversion ​

Author: Gordei Verbii
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.23094v1 Announce Type: cross Abstract: One implicit DDIM inversion step is the cheapest probe of whether a pretrained diffusion model encodes local manifold geometry. It is the stationarity condition of an explicit potential, $x-G(x)=\nabla\Psi_t(x)$, strongly convex at the Bayes limit wi...

📖 Read original article


271. Quantum Reservoir Computing with Physics-Informed Correction for Reduced-Order PDE Forecasting ​

Author: Krishna Bhatia, Harsh, Shalini Devendrababu
Published: 8/25/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.23119v1 Announce Type: cross Abstract: We study a hybrid proposal--correction architecture for reduced-order PDE forecasting in which a pure-state quantum reservoir computer (QRC) predicts latent coefficient dynamics and a PINN-based physics-informed corrector (PIC) refines local rollout ...

📖 Read original article


272. SGHA: A Single-Loop Fully First-Order Algorithm for Nonconvex-Strongly-Convex Bilevel Optimization ​

Author: Zhihao Gu, Qilong Wu, Junchi Yang
Published: 8/25/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2608.23211v1 Announce Type: cross Abstract: In this work, we study the oracle complexity of finding an $\epsilon$-stationary point for nonconvex-strongly-convex (NC-SC) bilevel optimization using only first-order oracles. Existing methods achieving the best-known complexity guarantees typicall...

📖 Read original article


273. BenthicDINO: Physics-Informed Self-Distillation for View-Invariant Side-Scan Sonar Representations ​

Author: Taqi Hamoda, Hayat Rajani, Nuno Gracias
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.23215v1 Announce Type: cross Abstract: Automated perception in side-scan sonar (SSS) imagery is severely hindered by physical acoustic artifacts, resulting in representations that inextricably mix intrinsic seabed reflectivity with transient viewing geometries. Existing self-supervised le...

📖 Read original article


274. Which Histories Matter for Time Series Forecasting? Learning Predictive Relevance with Future Supervision ​

Author: Yong-Hoon Choi, Youngjin Cho
Published: 8/25/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.23221v1 Announce Type: cross Abstract: Historical retrieval for time-series prediction commonly treats past similarity as a proxy for usefulness. We ask a different question: which historical examples should be expected to matter for a query? We define predictive relevance as expected fut...

📖 Read original article


275. Credal Large Language Models for Semantic Commitment under Uncertainty ​

Author: Shireen Kudukkil Manchingal, Sofiia Nikolenko, Fabio Cuzzolin
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, stat.ML

arXiv:2608.23244v1 Announce Type: cross Abstract: Large language models (LLMs) often produce fluent but incorrect answers with unwarranted confidence. A central limitation is that standard LLMs represent uncertainty through a single predictive distribution, conflating epistemic ignorance with genuin...

📖 Read original article


276. Apodex 1.1: Scaling Agentic Intelligence for Complex Work ​

Author: Apodex Team, B. An, B. Li, B. Wang, B. Zhang, B. L. Wang, C. Feng, C. Wei, C. Xue, C. Zhang, D. Ng, D. Ye, E. Min, F. Chen, F. Liu, F. Yang, F. Ye, H. Xu, H. Yang, H. Ye, H. Zhang, H. Zhao, J. Li, J. Lin, J. Xia, K. Jin, K. Wang, K. Yang, L. Bing, L. Lei, L. Su, Le. Wang, Lu. Wang, N. Wang, Q. Ren, Q. Yang, R. Li, S. Bai, S. Du, S. Li, S. Lin, S. Nie, S. Wang, S. Zhang, S. Z. Wang, Ta. Q. Fang, Ti. Q. Fang, W. Fang, W. Li, W. Zhang, X. Chen, X. Li, X. Tang, X. Wang, X. Xu, X. Zhang, X. Q. Wang, X. Y. Wang, Y. Deng, Y. Gao, Y. Hu, Y. Li, Y. Sui, Y. Wang, Y. Xiao, Y. Zhang, Z. Chen, Z. Cheng, Z. Feng, Z. Liang, Z. Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.23283v1 Announce Type: cross Abstract: General-purpose language models can reason and synthesize knowledge, but complex work also requires sustained interaction with files, information sources, and executable code, together with state maintenance, failure recovery, and verifiable delivery...

📖 Read original article


277. ADDA: a Modular Framework for Representing, Simulating and Assimilating Dynamics with End-to-end Differentiability ​

Author: Anthony Frion, Vien Minh Nguyen-Thanh, Ali Can Bekar, Pauleo R. Nimtz, Vadim Zinchenko, David S. Greenberg
Published: 8/25/2026, 4:00:00 AM
Categories: cs.MS, cs.LG

arXiv:2608.23297v1 Announce Type: cross Abstract: Data assimilation (DA) is an essential tool for prediction and understanding in the geosciences. DA combines simulation programs representing scientific knowledge with observations that constrain system dynamics, resulting in analyses and forecasts t...

📖 Read original article


278. Spicing up Genetic Netlist Generation with LLMs ​

Author: Stefan Uhlich, Ya\u{g}{\i}z Gen\c{c}er, Andrea Bonetti, Arun Venkitaraman, Chia-Yu Hsieh, Eisaku Ohbuchi, Lorenzo Servadei
Published: 8/25/2026, 4:00:00 AM
Categories: cs.NE, cs.AR, cs.LG

arXiv:2608.23317v1 Announce Type: cross Abstract: Analog circuit topology synthesis remains challenging because useful designs occupy a tiny fraction of a combinatorial search space, and small structural changes can induce highly nonlinear changes in behavior. Evolutionary algorithms are attractive ...

📖 Read original article


279. Beyond chlorophyll: machine learning estimates of diagnostic phytoplankton pigments from multispectral ocean colour data ​

Author: David Moffat, Angus Laurenson, Victor Martinez-Vicente, Gemma Kulk, Xuerong Sun, Robert J. W. Brewin, Shubha Sathyendranath
Published: 8/25/2026, 4:00:00 AM
Categories: q-bio.OT, cs.LG, physics.optics

arXiv:2608.23348v1 Announce Type: cross Abstract: Phytoplankton play a central role in marine ecosystems and the global carbon cycle, with different groups contributing differently to ocean biogeochemical processes. While standard techniques exist for monitoring phytoplankton concentration from ocea...

📖 Read original article


280. Hierarchical Exponential-Gaussian Mixtures for Watch-Time Distribution Prediction ​

Author: Sofia Gulevskaia, Mikhail Trapeznikov, Aleksandr Poslavsky, Alexander D'yakonov
Published: 8/25/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, stat.ML

arXiv:2608.23356v1 Announce Type: cross Abstract: Accurate watch-time (WT) prediction is an important requirement for short-video recommendations. Yet WT distributions are near-zero-inflated, long-tailed and multimodal. The recent Exponential-Gaussian Mixture Network (EGMN) models the full condition...

📖 Read original article


281. DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts ​

Author: Vlad Hondru, Florinel Alin Croitoru, Iuliana Georgescu, A. Sophia Koepke, Radu Tudor Ionescu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.23363v1 Announce Type: cross Abstract: Audio-visual deepfake detection is an actively studied topic, where one of the main challenges is to develop detectors able to generalize across deepfake generation methods. We conjecture that overfitting can be mitigated by extracting multiple high-...

📖 Read original article


282. KellyBoost: Growth-Optimal Portfolio Construction with Gradient-Boosted Trees ​

Author: Jiayu Li
Published: 8/25/2026, 4:00:00 AM
Categories: q-fin.PM, cs.LG

arXiv:2608.23393v1 Announce Type: cross Abstract: KellyBoost is a single multi-output XGBoost model whose softmax output is the portfolio: with y the vector of per-asset holding-period returns, the training loss is - log(1 + w y), the negative log growth rate, so the fitted model is the growth-optim...

📖 Read original article


283. Photorealistic Novel View Synthesis of Human Faces using Next-Scale Transformers ​

Author: Federico Stella, Fei Jiang, Zhongshi Jiang, Zohar Barzelay, Emanuel Garbin, Amin Jourabloo, Liuhao Ge
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.23410v1 Announce Type: cross Abstract: Photorealistic novel view synthesis of people remains challenging at high spatial resolutions and across multiple target cameras, where preserving identity, fine appearance details, and geometric coherence is critical. We build on the next-scale auto...

📖 Read original article


284. Exploring Long-period Architectures: Four New Planet Candidates from Kepler with Periods >342 days ​

Author: Matthew T. Hansen, Jason A. Dittmann
Published: 8/25/2026, 4:00:00 AM
Categories: astro-ph.EP, astro-ph.IM, cs.LG

arXiv:2608.23425v1 Announce Type: cross Abstract: The Kepler detection pipeline, as well as the transit method, has a bias towards shorter periods, leaving a dearth of detections at longer orbital periods. This relative lack of detections has left an incomplete picture of the architectures of exopla...

📖 Read original article


285. Reward-Free Continual Adaptation for Resilient Space Robots ​

Author: Andrej Orsula, Miguel Olivares-Mendez, Carol Martinez
Published: 8/25/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2608.23452v1 Announce Type: cross Abstract: Space robots operate in extreme environments where hardware degradation can critically compromise traditional control strategies. While continual reinforcement learning offers a promising mechanism for online adaptation, it inherently requires access...

📖 Read original article


286. Primal--Dual Alternating Neural Learning for Timely Classification with Performance Guarantees ​

Author: Jiaming Qiu, Yingye Zheng, Ying-Qi Zhao
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.23480v1 Announce Type: cross Abstract: Timely risk classification is essential in many clinical monitoring settings, where decisions must balance the benefit of classifying patients early for subsequent intervention against the value of observing additional data. Yet most existing statist...

📖 Read original article


Author: Santosh Ray, Pratik K. Mishra, Ali Abedi, Charlene H. Chu, Amir Ahmad, Shehroz S. Khan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.23531v1 Announce Type: cross Abstract: Older adults recovering after lower-limb fracture or hip replacement may experience complex recovery trajectories. Most of the time, these clinical aspects are studied in isolation, masking their joint impact on recovery. This study used the MAISON-L...

📖 Read original article


288. Interpretable AI with Local Distillation ​

Author: Erin Craig, Yiling Huang, Snigdha Panigrahi
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML

arXiv:2608.23538v1 Announce Type: cross Abstract: Modern AI models such as tabular foundation models and gradient-boosted ensembles can outpredict classical methods, but provide little basis for reasoning about their predictions. High-stakes decisions call for models that are both accurate and inter...

📖 Read original article


289. Inertial Manifold Neural Operator for Dissipative Time-Dependent Partial Differential Equations ​

Author: Xiaoyang Xie, Clarence W. Rowley
Published: 8/25/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.DS

arXiv:2608.23546v1 Announce Type: cross Abstract: In this paper, we introduce the Inertial Manifold Neural Operator (IMNO) for solving dissipative time-dependent partial differential equations (PDEs). The long-time dynamics of such systems often exhibit an effective low-dimensional structure due to ...

📖 Read original article


290. Robustness of Anomaly Detection Models for Industrial Control Systems under Training-Time Data Contamination ​

Author: Mustafa Umut Ozbek, Taiwo Ojo, Pooria Madani, Khalil El-Khatib, Li Yang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.23547v1 Announce Type: cross Abstract: Machine-learning-based anomaly detection is increasingly used in industrial control systems (ICS), yet most studies assume that detector training data is trustworthy. In practice, training data may be corrupted through compromised logs, labeling erro...

📖 Read original article


291. ConvergeFlow: Language Flow with Provable Convergence to Token Embeddings ​

Author: Na Li, Yuchen Jiao, Changxiao Cai, Gen Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, stat.ML

arXiv:2608.23551v1 Announce Type: cross Abstract: Recent advances in continuous diffusion and flow-based language models (LMs) have achieved performance competitive with discrete LMs. However, existing continuous frameworks still rely on decoders supervised with cross entropy (CE) because the flow t...

📖 Read original article


292. Residual-based attention in physics-informed neural networks ​

Author: Sokratis J. Anagnostopoulos, Juan Diego Toscano, Nikolaos Stergiopulos, George Em Karniadakis
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph

arXiv:2307.00379v2 Announce Type: replace Abstract: Driven by the need for more efficient and seamless integration of physical models and data, physics-informed neural networks (PINNs) have seen a surge of interest in recent years. However, ensuring the reliability of their convergence and accuracy ...

📖 Read original article


293. Learning to Select and Rank from Choice-Based Feedback: A Simple Nested Approach ​

Author: Junwen Yang, Yifan Feng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2307.09295v3 Announce Type: replace Abstract: We study a ranking and selection problem of learning from choice-based feedback with dynamic assortments. In this problem, a company sequentially displays a set of items to a population of customers and collects their choices as feedback. The only ...

📖 Read original article


294. Learning in PINNs: Phase transition, diffusion equilibrium, and generalization ​

Author: Sokratis J. Anagnostopoulos, Juan Diego Toscano, Nikolaos Stergiopulos, George Em Karniadakis
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2403.18494v2 Announce Type: replace Abstract: We investigate the learning dynamics of fully-connected neural networks through the lens of the neural gradient signal-to-noise ratio (SNR), examining the behavior of first-order optimizers in non-convex objectives. Interpreting the drift/diffusion...

📖 Read original article


295. Soft Label PU Learning ​

Author: Yong Lei, Puning Zhao, Jintao Deng, Xu Cheng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2405.01990v2 Announce Type: replace Abstract: PU learning refers to the classification problem in which only part of positive samples are labeled. Existing PU learning methods treat unlabeled samples equally. However, in many real tasks, from common sense or domain knowledge, some unlabeled sa...

📖 Read original article


296. DuoGNN: Topology-aware Graph Neural Network with Homophily and Heterophily Interaction-Decoupling ​

Author: K. Mancini, I. Rekik
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.SI

arXiv:2409.19616v3 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have proven effective in various medical imaging applications, such as automated disease diagnosis. However, due to the local neighborhood aggregation paradigm in message passing which characterizes these models, they i...

📖 Read original article


297. Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations ​

Author: Yonatan Sverdlov, Ido Springer, Nadav Dym
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2410.06665v5 Announce Type: replace Abstract: This paper explores the characterization of equivariant linear layers for representations of permutations and related groups. Unlike traditional approaches, which address these problems using parameter-sharing, we consider an alternative methodolog...

📖 Read original article


298. Learning Mamba as a Continual Learner: Meta-Learning Selective State Space Models for Continual Learning ​

Author: Chongyang Zhao, Dong Gong
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2412.00776v5 Announce Type: replace Abstract: Continual learning (CL) learns from a non-stationary data stream without storing or re-training on all seen samples. Meta-continual learning (MCL) casts CL as sequence prediction and meta-learns the continual learner itself as a sequence model, wit...

📖 Read original article


299. NeST: Neighborhood-aware semantic alignment and temporal modulation for LLM based time series forecasting ​

Author: Jayanie Bogahawatte, Sachith Seneviratne, Maneesha Perera, Saman Halgamuge
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2412.04806v2 Announce Type: replace Abstract: Adapting Large Language Models (LLMs) trained on discrete text data, to forecast continuous time series signals is challenging. While finetuning the LLMs enables such adaptation, effectively integrating both textual and time series information in t...

📖 Read original article


300. DeltaGNN: Graph Neural Network with Information Flow Control ​

Author: Kevin Mancini, Islem Rekik
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2501.06002v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are popular deep learning models designed to process graph-structured data through recursive neighborhood aggregations in the message passing process. When applied to semi-supervised node classification, the message-pas...

📖 Read original article


301. SRMT: Shared Memory for Multi-agent Lifelong Pathfinding ​

Author: Alsu Sagirova, Yuri Kuratov, Mikhail Burtsev
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA

arXiv:2501.13200v2 Announce Type: replace Abstract: Coordination in decentralized multi-agent reinforcement learning (MARL) necessitates that agents share information about their behavior and intentions. Existing approaches rely on communication protocols with domain or resource constraints or centr...

📖 Read original article


302. Fairness-Aware Low-Rank Representation Fine-Tuning ​

Author: Parameswaran Kamalaruban, Mark Anderson, Stuart Burrell, Maeve Madigan, Piotr Skalski, David Sutton
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2503.05684v2 Announce Type: replace Abstract: Pre-trained foundation models can be efficiently adapted for specific tasks using Low-Rank Adaptation (LoRA), but the fairness properties of these adapted classifiers remain underexplored. Existing fairness-aware fine-tuning methods assume that sen...

📖 Read original article


303. Two Stage Wireless Federated LoRA Fine-Tuning with Sparsified Orthogonal Updates ​

Author: Bumjun Kim, Wan Choi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2505.00333v3 Announce Type: replace Abstract: Federated fine-tuning with low-rank adaptation (LoRA) communicates only two low-rank matrices instead of the full model, but existing methods typically fix the LoRA rank in advance as a manually tuned hyperparameter. In wireless networks, however, ...

📖 Read original article


304. Learning with Local Search MCMC Layers ​

Author: Germain Vivier-Ardisson, Mathieu Blondel, Axel Parmentier
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2505.14240v3 Announce Type: replace Abstract: Integrating combinatorial optimization layers into neural networks has recently attracted significant research interest. However, many existing approaches lack theoretical guarantees or fail to perform adequately when relying on inexact solvers. Th...

📖 Read original article


305. Model Merging is Secretly Certifiable: Non-Vacuous Generalisation Bounds for Low-Shot Learning ​

Author: Taehoon Kim, Henry Gouk, Minyoung Kim, Timothy Hospedales
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2505.15798v2 Announce Type: replace Abstract: Certifying the IID generalisation ability of deep networks is the first of many requirements for trusting AI in high-stakes applications from medicine to security. However, when instantiating generalisation bounds for deep networks it remains chall...

📖 Read original article


306. VIBE: Vector Index Benchmark for Embeddings ​

Author: Elias J"a"asaari, Ville Hyv"onen, Matteo Ceccarello, Teemu Roos, Martin Aum"uller
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2505.17810v3 Announce Type: replace Abstract: Approximate nearest neighbor (ANN) search is a performance-critical component of many machine learning pipelines, and rigorous benchmarking is essential for assessing the performance of vector indexes for ANN search. However, the datasets of existi...

📖 Read original article


307. One-shot Robust Federated Learning of Independent Component Analysis ​

Author: Dian Jin, Xin Bing, Yuqian Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2505.20532v3 Announce Type: replace Abstract: This paper studies robust one-shot aggregation for distributed and federated Independent Component Analysis (ICA). In this setting, each client computes a local ICA estimator, while the server aims to recover a common global mixing matrix without a...

📖 Read original article


308. Attribute-Efficient PAC Learning of Sparse Halfspaces with Constant Malicious Noise Rate ​

Author: Shiwei Zeng, Jie Shen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2505.21430v3 Announce Type: replace Abstract: Attribute-efficient PAC learning of sparse halfspaces has been a fundamental problem in machine learning theory. In recent years, machine learning algorithms are faced with prevalent data corruptions or even malicious attacks. It is of central inte...

📖 Read original article


309. Scaling Electronic Health Record Foundation Models for Population Health Management ​

Author: Liwen Sun, Hao-Ren Yao, Ophir Frieder, Xiang Qian, Chenyan Xiong
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2506.00209v3 Announce Type: replace Abstract: Population health management requires scalable methods to identify individuals at risk of chronic diseases such as cardiovascular conditions and cancer, yet existing approaches rely on fragmented data and resource-intensive screening. We present Sc...

📖 Read original article


310. Time Series Forecasting via Reasoning: A Slow-Thinking Approach with Reinforcement Fine-Tuned LLMs ​

Author: Yitong Zhou, Yucong Luo, Mingyue Cheng, Qi Liu, Jiahao Wang, Daoyu Wang, Enhong Chen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2506.10630v4 Announce Type: replace Abstract: To advance time series forecasting (TSF), various methods have been proposed to improve prediction accuracy, evolving from statistical techniques to data-driven deep learning architectures. Despite their effectiveness, most existing methods still a...

📖 Read original article


311. Seismic Acoustic Impedance Inversion Framework Based on Conditional Latent Generative Diffusion Model ​

Author: Jie Chen, Hongling Chen, Jinghuai Gao, Chuangji Meng, Tao Yang, XinXin Liang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2506.13529v2 Announce Type: replace Abstract: Seismic acoustic impedance plays a crucial role in lithological identification and subsurface structure interpretation. However, due to the inherently ill-posed nature of the inversion problem, directly estimating impedance from post-stack seismic ...

📖 Read original article


312. Persistent Homology as a Theory of Emergent Structure ​

Author: Xin Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2507.03065v3 Announce Type: replace Abstract: Why do some macroscopic structures remain identifiable even though their microscopic constituents continually change? Vortices persist while fluid parcels turn over, neural memories persist while spikes and synapses fluctuate, and institutions pers...

📖 Read original article


313. Adversarial Training Improves Generalization Under Distribution Shifts in Bird Sound Classification ​

Author: Ren'e Heinrich, Lukas Rauch, Raphael Schwinger, Katharina Brauns, Bastian Sch"afermeier, Bernhard Sick, Christoph Scholz
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2507.13727v2 Announce Type: replace Abstract: Adversarial training is a promising strategy for enhancing robustness against adversarial attacks, but its impact on generalization under substantial distribution shifts in audio classification remains largely unexplored. We address this gap by inv...

📖 Read original article


314. Entity Representation Learning Through Onsite-Offsite Graph for Pinterest Ads ​

Author: Jiayin Jin, Erika Sun, Zhimeng Pan, Yang Tang, Jiarui Feng, Kungang Li, Chongyuan Xiang, Jiacheng Li, Runze Su, Siping Ji, Han Sun, Ling Leng, Prathibha Deshikachar
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE

arXiv:2508.02609v3 Announce Type: replace Abstract: Graph Neural Networks (GNN) have been extensively applied to industry recommendation systems, as seen in models like GraphSage\cite{GraphSage}, TwHIM\cite{TwHIM}, LiGNN\cite{LiGNN} etc. In these works, graphs were constructed based on users' activi...

📖 Read original article


315. Bidding-Aware Retrieval for Multi-Stage Consistency in Online Advertising ​

Author: Bin Liu, Yunfei Liu, Ziru Xu, Zhi Kou, Yeqiu Yang, Han Zhu, Jian Xu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2508.05206v2 Announce Type: replace Abstract: Online advertising systems typically use a cascaded architecture to manage massive requests and candidate volumes, where the ranking stages allocate traffic based on eCPM (predicted CTR $\times$ Bid). With the increasing popularity of auto-bidding ...

📖 Read original article


316. LISM: Long-range Integrative State space Models via Input-Latent State Interactions ​

Author: Cong Ma, Kayvan Najarian, Hovhannes Baghdasaryan, Hrachya Astsatryan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.04226v2 Announce Type: replace Abstract: State space models (SSMs) are an emerging paradigm that achieves linear-time scaling, however, they intrinsically suffer from "curse of memory". The memory of SSMs, including Mamba, decays exponentially as long as the recursive update is stable. In...

📖 Read original article


317. FedRP: A Communication-Efficient Approach for Differentially Private Federated Learning Using Random Projection ​

Author: Sina Najafi, Mostafa Tavassolipour, Mohammad Hasan Narimani
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.10041v2 Announce Type: replace Abstract: Federated learning (FL) enables collaborative model training without centralizing data, but exchanging high-dimensional updates can expose sensitive information and incur substantial communication costs. We present FedRP, a communication-efficient ...

📖 Read original article


318. HD3C: Efficient Medical Data Classification for Edge Devices ​

Author: Jianglan Wei, Zhenyu Zhang, Pengcheng Wang, Mingjie Zeng, Zhigang Zeng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.14617v5 Announce Type: replace Abstract: Efficient medical data classification is essential for modern disease screening, particularly in resource-constrained environments where power budgets and computing capabilities are limited. We present HD3C, a lightweight classification framework d...

📖 Read original article


319. HiViS: Hiding Visual Tokens from the Drafter for Speculative Decoding in Vision-Language Models ​

Author: Zhinan Xie, Peisong Wang, Shuang Qiu, Jian Cheng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2509.23928v3 Announce Type: replace Abstract: Speculative decoding has proven effective for accelerating inference in Large Language Models (LLMs), yet its extension to Vision-Language Models (VLMs) remains limited by the computational burden and semantic inconsistency introduced by visual tok...

📖 Read original article


320. SLogic: Subgraph-Informed Logical Rule Learning for Knowledge Graph Completion ​

Author: Trung Hoang Le, Tran Cao Son, Ishtiaq Ahmed, Huiping Cao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.00279v3 Announce Type: replace Abstract: Logical rule-based methods offer an interpretable approach to knowledge graph completion (KGC) by capturing compositional relationships in the form of human-readable inference rules. While existing logical rule-based methods learn rule confidence s...

📖 Read original article


321. RSTGCN: Railway-centric Spatio-Temporal Graph Convolutional Network for Train Delay Prediction ​

Author: Koyena Chowdhury, Paramita Koley, Abhijnan Chakraborty, Saptarshi Ghosh
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.01262v2 Announce Type: replace Abstract: Accurate prediction of train delays is critical for efficient railway operations. While earlier approaches have largely focused on forecasting the exact delays of individual trains, studies on station-level delay prediction are somewhat sparse. To ...

📖 Read original article


322. R\'enyi Sharpness: A Novel Sharpness that Strongly Correlates with Generalization ​

Author: Qiaozhe Zhang, Jun Sun, Ruijie Zhang, Yingzhuang Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.07758v4 Announce Type: replace Abstract: Sharpness (of the loss minima) is widely believed to be a good indicator of generalization of neural networks. Unfortunately, the correlation between existing sharpness measures and generalization is not as strong as expected, and sometimes even co...

📖 Read original article


323. Closing the Curvature Gap: Full Transformer Hessians ​

Author: Egor Petrov, Nikita Kiselev, Vladislav Meshkov, Andrey Grabovoy
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.16927v2 Announce Type: replace Abstract: The optimization landscape of Transformer models remains poorly understood despite their widespread adoption. While recent studies have derived curvature properties for isolated self-attention mechanisms, a comprehensive theoretical characterizatio...

📖 Read original article


324. Forgetting to Forget: Attention Sink as A Gateway for Backdooring LLM Unlearning ​

Author: Bingqi Shang, Yiwei Chen, Yihua Zhang, Bingquan Shen, Sijia Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2510.17021v2 Announce Type: replace Abstract: Large language model (LLM) unlearning is a key approach for removing undesired data, knowledge, or behaviors from pretrained models while retaining their general utility. Yet, with the rise of open-weight LLMs, we ask: can the unlearning process it...

📖 Read original article


325. A New Type of Adversarial Examples ​

Author: Xingyang Nie, Caoliang Zhang, Su Pan, Biao Wang, Huilin Ge, Tao Fang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GR

arXiv:2510.19347v2 Announce Type: replace Abstract: Most machine learning models are vulnerable to adversarial examples, which poses security concerns on these models. Adversarial examples are crafted by applying subtle but intentionally worst-case modifications to examples from the dataset, leading...

📖 Read original article


326. Mitigating Sample-Level Imbalance via Probabilistic Separation for Adaptive Multimodal Fusion ​

Author: Zhiwen Yu, Zhaocheng Liu, Xiaoqing Liu, Huanqiang Zeng, C. L. Philip Chen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SD, eess.AS

arXiv:2510.21797v4 Announce Type: replace Abstract: Multimodal learning faces modality imbalance, where dominant modalities suppress weaker ones due to inconsistent convergence rates. Existing static or heuristic methods overlook sample-level variations in prediction bias and fail to isolate low-qua...

📖 Read original article


327. Decomposable Neural Symbolic Regression ​

Author: Giorgio Morales, John W. Sheppard
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.04124v4 Announce Type: replace Abstract: Symbolic regression (SR) models complex systems by discovering mathematical expressions that capture underlying relationships in observed data. However, most SR methods prioritize minimizing prediction error over identifying the governing equations...

📖 Read original article


328. Radial Compensation: The Inverse Base-Distribution Problem for Chart-Based Generative Models on Riemannian Manifolds ​

Author: Marios Papamichalis, Regina Ruane
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.DG, math.IT, stat.ML

arXiv:2511.14056v3 Announce Type: replace Abstract: Latent-variable models on spheres and hyperbolic spaces usually draw a Gaussian in the tangent space at a base point and push it onto the manifold. On these spaces the distance from the base point is the coordinate that carries meaning: depth in a ...

📖 Read original article


329. Finite-Nudge Equilibrium Propagation in Thermal Ensembles ​

Author: Elon Litman
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2511.22024v2 Announce Type: replace Abstract: We liberate Equilibrium Propagation (EP) from the limit of infinitesimal perturbations by establishing a finite-nudge foundation for local credit assignment. By modeling network states as Gibbs-Boltzmann distributions rather than deterministic poin...

📖 Read original article


330. Directed evolution algorithm drives neural prediction ​

Author: Yanlin Wang, Nancy M Young, Patrick C M Wong
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.01362v3 Announce Type: replace Abstract: Neural prediction offers a promising approach to forecasting the individual variability of neurocognitive functions and disorders and providing prognostic indicators for personalized invention. However, it is challenging to translate neural predict...

📖 Read original article


331. Diagnosing Capability Preservation and Task Sensitivity in Memory Augmented Document Classifiers ​

Author: Isaac Kofi Nti
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE

arXiv:2512.06582v2 Announce Type: replace Abstract: End task accuracy alone cannot determine whether a memory mechanism preserves an acquired capability, exposes sample-specific stored information, or contributes measurably to downstream performance. This study introduces Protected QL Memory and eva...

📖 Read original article


332. Chorus: Harmonizing Context and Sensing Signals for Data-Free Model Customization in IoT ​

Author: Liyu Zhang, Yejia Liu, Kwun Ho Liu, Runxi Huang, Xiaomin Ouyang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.15206v3 Announce Type: replace Abstract: A key bottleneck toward scalable IoT sensing is efficiently adapting trained AI models to new deployment conditions. Context shifts, such as changes in sensor placement or ambient environments, can substantially alter sensing patterns and degrade m...

📖 Read original article


333. You Need Better Attention Priors ​

Author: Elon Litman, Gabe Guo
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, stat.ML

arXiv:2601.15380v2 Announce Type: replace Abstract: We generalize the attention mechanism by viewing it through the lens of Entropic Optimal Transport, revealing that standard attention corresponds to a transport problem regularized by an implicit uniform prior. We introduce Generalized Optimal tran...

📖 Read original article


334. MADE: Benchmark Environments for Closed-Loop Materials Discovery ​

Author: Shreshth A Malik, Tiarnan Doherty, Panagiotis Tigas, Muhammed Razzak, Stephen J. Roberts, Aron Walsh, Yarin Gal
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci

arXiv:2601.20996v2 Announce Type: replace Abstract: Existing benchmarks for computational materials discovery primarily evaluate static predictive tasks or isolated computational sub-tasks. While valuable, these evaluations neglect the inherently iterative and adaptive nature of scientific discovery...

📖 Read original article


335. Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks ​

Author: Luwei Sun, Dongrui Shen, Feng Chuanwen, Jianfe Li, Yulong Zhao, Han Feng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.21242v2 Announce Type: replace Abstract: Motivated by challenges in conditional generative modeling, where the target conditional density takes the form of a ratio f1 over f2, this paper develops a theoretical framework for approximating such ratio-type functionals. Here, f1 and f2 are ke...

📖 Read original article


336. Mode-Dependent Rectification for Stable PPO Training ​

Author: Mohamad Mohamad, Francesco Ponzio, Xavier Descombes
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.05619v2 Announce Type: replace Abstract: Mode-dependent architectural components (layers that behave differently during training and evaluation, such as Batch Normalization or dropout) are commonly used in visual reinforcement learning but can destabilize on-policy optimization. We show t...

📖 Read original article


337. Which Algorithms Can Graph Neural Networks Learn? ​

Author: Solveig Wittig, Antonis Vasileiou, Robert R. Nerem, Timo Stoll, Floris Geerts, Yusu Wang, Christopher Morris
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DS, cs.NE

arXiv:2602.13106v2 Announce Type: replace Abstract: In recent years, there has been growing interest in understanding neural architectures' ability to learn to execute discrete algorithms, a line of work often referred to as neural algorithmic reasoning. The goal is to integrate algorithmic reasonin...

📖 Read original article


338. Optimizer choice matters for the emergence of Neural Collapse ​

Author: Jim Zhao, Tin Sum Cheng, Wojciech Masarczyk, Aurelien Lucchi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.16642v4 Announce Type: replace Abstract: Neural Collapse (NC) refers to the emergence of highly symmetric geometric structures in the representations of deep neural networks during the terminal phase of training. Despite its prevalence, the theoretical understanding of NC remains limited....

📖 Read original article


339. Continual Uncertainty Learning for Robust Control of Nonlinear Systems with Multiple Heterogeneous Uncertainties ​

Author: Heisei Yonezawa, Ansei Yonezawa, Itsuro Kajiwara
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY

arXiv:2602.17174v3 Announce Type: replace Abstract: Robust control of mechanical systems with multiple uncertainties remains a fundamental challenge, particularly when nonlinear dynamics and operating-condition variations are intricately intertwined. Although deep reinforcement learning combined wit...

📖 Read original article


340. Learning with Boolean threshold functions ​

Author: Veit Elser, Manish Krishan Lal
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17493v2 Announce Type: replace Abstract: We develop a method for training neural networks on Boolean data in which the values at all nodes are strictly $\pm 1$, and the resulting models are typically equivalent to networks whose nonzero weights are also $\pm 1$. The method replaces loss m...

📖 Read original article


341. The Error of Deep Operator Networks Is the Sum of Its Parts: Branch-Trunk and Mode Error Decompositions ​

Author: Alexander Heinlein, Johannes Taraz
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2602.21910v2 Announce Type: replace Abstract: Operator learning has the potential to strongly impact scientific computing by learning solution operators for differential equations, potentially accelerating multi-query tasks such as design optimization and uncertainty quantification by orders o...

📖 Read original article


342. On the Convergence of Single-Loop Stochastic Bilevel Optimization with Approximate Implicit Differentiation ​

Author: Yubo Zhou, Luo Luo, Guang Dai, Haishan Ye
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.23633v2 Announce Type: replace Abstract: Stochastic Bilevel Optimization has emerged as a fundamental framework for meta-learning and hyperparameter optimization. Despite the practical prevalence of single-loop algorithms, their theoretical understanding in the stochastic regime remains l...

📖 Read original article


343. Safety Training May Persist Through Helpfulness Optimization in LLM Agents ​

Author: Benjamin Plaut
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2603.02229v2 Announce Type: replace Abstract: Safety post-training has been studied extensively in single-step "chat" settings where safety typically refers to refusing harmful requests. We study an "agentic" (i.e., multi-step, tool-use) setting where safety refers to harmful actions directly ...

📖 Read original article


344. PRAGMA: Revolut Foundation Model ​

Author: Maxim Ostroukhov, Ruslan Mikhailov, Vladimir Iashin, Artem Sokolov, Andrei Akshonov, Vitaly Protasov, Andrey Goncharov, Dmitrii Beloborodov, Vince Mullin, Roman Yokunda Enzmann, Georgios Kolovos, Jason Renders, Pavel Nesterov, Anton Repushko
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.CL, cs.IR, q-fin.CP

arXiv:2604.08649v2 Announce Type: replace Abstract: Modern financial systems generate vast quantities of transactional and event-level data that encode rich economic signals. This paper presents PRAGMA, a family of foundation models for banking event sequences. Our approach pre-trains a Transformer-...

📖 Read original article


345. When Less Latent Leads to Better Relay: Information-Preserving Compression for Latent Multi-Agent LLM Collaboration ​

Author: Yiping Li, Zhiyu An, Wan Du
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.13349v3 Announce Type: replace Abstract: Multi-agent LLM systems are moving beyond discrete-token messages toward richer relays that preserve internal state. Recent work such as LatentMAS transmits full key-value (KV) caches between agents but pays a high memory and communication cost. We...

📖 Read original article


346. Detecting and Suppressing Reward Hacking with Gradient Fingerprints ​

Author: Songtao Wang, Quang Hieu Pham, Fangcong Yin, Xinpeng Wang, Jocelyn Qiaochu Chen, Greg Durrett, Xi Ye
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2604.16242v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) typically optimizes for outcome rewards without imposing constraints on intermediate reasoning. This leaves training susceptible to reward hacking, where models exploit loopholes (e.g., spurious...

📖 Read original article


347. FSEVAL: Feature Selection Evaluation Toolbox and Dashboard ​

Author: Muhammad Rajabinasab, Arthur Zimek
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.18227v4 Announce Type: replace Abstract: Feature selection is a fundamental machine learning and data mining task, involved with discriminating redundant features from informative ones. It is an attempt to address the curse of dimensionality by removing the redundant features, while unlik...

📖 Read original article


348. Collocation-based Robust Physics Informed Neural Networks for time-dependent simulations of pollution propagation under thermal inversion conditions on Spitsbergen ​

Author: Maciej Sikora, Leszek Siwik, Natalia Leszczy'nska, Tomasz Maciej Ciesielski, Eirik Valseth, Manuela Bastidas Olivares, Marcin {\L}o's, Tomasz S{\l}u.zalec, Jacek Leszczy'nski, Maciej Paszy'nski
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2604.23003v2 Announce Type: replace Abstract: In this paper, we propose a Physics-Informed Neural Network framework for time-dependent simulations of pollution propagation originating from moving emission sources. We formulate a robust variational framework for the time-dependent advection-dif...

📖 Read original article


349. DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Discrete Diffusion Models ​

Author: Dake Bu, Wei Huang, Andi Han, Si Wu, Hau-San Wong, Qingfu Zhang, Taiji Suzuki, Atsushi Nitanda
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.24357v3 Announce Type: replace Abstract: Discrete diffusion models admit many token orders, yet most systems rely on confidence-based decoding. Confidence is a strong and efficient heuristic, but it can be myopic because local certainty does not measure a position's effect on terminal qua...

📖 Read original article


350. DR-SNE: Density-Regularized Stochastic Neighbor Embedding ​

Author: Maksim Kazanskii
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.02060v2 Announce Type: replace Abstract: Dimensionality-reduction methods such as t-SNE preserve local neighborhood structure but can substantially distort the local distribution of data. We introduce Density-Regularized Stochastic Neighbor Embedding (DR-SNE), which augments stochastic ne...

📖 Read original article


351. GraphSVR: A Graph Convolutional Support Vector Regression Framework for Robust Spatiotemporal Air Pollution Forecasting ​

Author: Nourin Jahan, Muhammed Navas T, Tanujit Chakraborty, Madhurima Panja
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, stat.AP, stat.ML

arXiv:2605.03795v3 Announce Type: replace Abstract: Urban air quality forecasting is challenging because pollutant concentrations are nonlinear, nonstationary, spatiotemporally dependent, and often affected by anomalous observations caused by traffic congestion, industrial emissions, and seasonal me...

📖 Read original article


352. Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks ​

Author: Pavel Prochazka
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.08446v5 Announce Type: replace Abstract: The standard training objectives of Bayesian deep learning are posterior-seeking: their optimum over the belief is the posterior of a fitted model, or its KL projection. We show that the shared target is a removable constraint on the belief-not an ...

📖 Read original article


353. System-Prompt Anchoring with Cross-Attention Layers ​

Author: Li Lixing
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.09737v2 Announce Type: replace Abstract: Cross-attention provides a dedicated route from a selected information source into a model's computation, but the effect of where that route is inserted remains underexplored. We study this question when the source is a privileged system-prompt spa...

📖 Read original article


354. Unlocking air traffic flow prediction through microscopic aircraft-state modeling ​

Author: Bin Wang, Anqi Liu, Jiangtao Zhao, Yanyong Huang, Hina Birahmani, Peilan He, Guiyuan Jiang, Feng Hong, Yanwei Yu, Yuanyuan Hou, Tianrui Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.10083v3 Announce Type: replace Abstract: Short-term air traffic flow prediction in terminal airspace is essential for proactive air traffic management. Existing approaches predominantly model traffic flow as aggregated time series. However, traffic dynamics are governed by aircraft states...

📖 Read original article


355. Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective ​

Author: Feng Zhang, Xinhong Ma, Ziqiang Dong, Xi Leng, Jianfei Zhao, Xin Sun, Yang Yang, Guanjun Jiang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.12969v4 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) is one of the most widely adopted RLVR algorithms for post-training large language models on reasoning tasks. We first show that GRPO admits an equivalent discriminative reformulation, in which policy optim...

📖 Read original article


356. Edge-AI-Driven Learning-to-Rank for Decentralized Task Allocation in Circular Smart Manufacturing ​

Author: Mohammadhossein Ghahramani, Yan Qiao, Mengchu Zhou
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.16433v2 Announce Type: replace Abstract: Task allocation in smart manufacturing systems must operate under decentralized decision-making, dynamic workloads, and shared-resource constraints. In circular manufacturing settings, these challenges are further intensified because tasks compete ...

📖 Read original article


357. Boundedly Rational Meta-Learning in Sequential Consumer Choice ​

Author: Mehrzad Khosravi, Max Kleiman-Weiner, Hema Yoganarasimhan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, econ.GN, q-fin.EC

arXiv:2605.16532v2 Announce Type: replace Abstract: Many consumer decisions involve repeated choices under uncertainty, where experience in one context may inform decisions in another. For example, experience with a brand in one market or usage context may shape beliefs about that brand in a new con...

📖 Read original article


358. Learning What Evaluators Value: A Reliable Approach to Modeling Evaluator Preferences ​

Author: Madeline Celi Kitch, Nihar B. Shah
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.16615v2 Announce Type: replace Abstract: In many applications, human and LLM evaluators use assessments of relevant criteria to create an overall evaluation for an item or individual. For example, in admissions, committees assess candidates on attributes such as test scores, GPA, and rese...

📖 Read original article


359. Localize and Neutralize: Gradient-guided Token Suppression against Visual Prompt Injection Attack ​

Author: Dongpeng Zhang, Ke Ma, Yangbangyan Jiang, Gaozheng Pei, Longtao Huang, Qianqian Xu, Qingming Huang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.25194v3 Announce Type: replace Abstract: Adversarial images pose a severe security threat to multimodal large language models through prompt injection. Existing defenses largely lack a principled understanding of the underlying mechanisms and struggle to balance efficiency and defense uti...

📖 Read original article


360. How Should LLMs Consume High-Quality Data? Optimal Data Scheduling via Quality-Aware Functional Scaling Laws ​

Author: Zhitao Zhu, Xili Wang, Shizhe Wu, Jiawei Fu, Xiaoqing Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.25698v2 Announce Type: replace Abstract: High-quality data is scarce in large language model (LLM) training, yet how to schedule its use with optimization dynamics lacks theoretical guidance. We extend functional scaling laws with time-varying data quality and derive asymptotically optima...

📖 Read original article


361. BIRDNet: Mining and Encoding Boolean Implication Knowledge Graphs as Interpretable Deep Neural Networks ​

Author: Tirtharaj Dash
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE, q-bio.QM

arXiv:2605.28739v2 Announce Type: replace Abstract: Tabular data in knowledge-rich domains often carries a latent prior in the form of Boolean implication relationships (BIRs) between pairs of features. We mine such relationships with a sparse-exception binomial test. We encode the resulting typed g...

📖 Read original article


362. Same Payload, Different Channel: Measuring Trust Asymmetry in Tool-Using Language Models ​

Author: Mohammed Sameer Syed (University of Arizona), Rozhin Yasaei (University of Arizona)
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CR

arXiv:2606.00566v2 Announce Type: replace Abstract: As language models take on agentic roles that call APIs, read tool outputs, and act on third-party content, their attack surface expands beyond what users type. Whether they treat a malicious instruction the same way regardless of where it arrives ...

📖 Read original article


363. Enhancing LLM Metacognition via Cognitive Pairwise Training ​

Author: Weitao Li, Hao Zhou, Xuanyu Lei, Fandong Meng, Yuanhang Liu, Jingyi Ren, Ante Wang, Xiaolong Wang, Yuanchi Zhang, Fuwen Luo, Guangwen Yang, Lin Gan, Weizhi Ma, Yang Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.00869v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become central to LLM reasoning, but its outcome-level rewards can make models more willing to give confident answers when evidence or reasoning is unreliable. Existing SFT or RL methods mai...

📖 Read original article


364. Soft-NBCE: Entropy-Weighted Chunk Fusion for Long-Context ​

Author: Shihao Ji, Mingyu Li, Zihui Song
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.01101v2 Announce Type: replace Abstract: The quadratic complexity of self-attention remains a bottleneck for Large Language Models (LLMs) processing ultra-long contexts. The Naive Bayes Cognitive Engine (NBCE) parallelizes long-context inference by chunking documents and routing to the lo...

📖 Read original article


365. Shortcomings and capacities of real-constrained neural networks in complex spaces ​

Author: Andrew Gracyk
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, math.PR

arXiv:2606.04390v2 Announce Type: replace Abstract: We find the asymptotic ratio between the storage capacities when enforcing real pre-activations in a complex hypothesis class as opposed to complex ones in the same class. Our methods depend on Gardner volume comparisons at critical capacity. Our p...

📖 Read original article


366. From Cone Geometry to Monge Structure: Local Defects and Global Regret in High-Dimensional Optimal Transport ​

Author: Lei Luo, Jian Yang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.04695v2 Announce Type: replace Abstract: Exact optimal transport in high dimensions becomes tractable only when additional structure selects the optimal coupling. Classical Monge transportation solves the discrete problem once a cost array is known to be Monge, but it does not explain whe...

📖 Read original article


367. Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning ​

Author: Yuhuan Yuan, Zhouliang Yu, Minghao Liu, Weiyang Liu, Ge Lin Kan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.07602v2 Announce Type: replace Abstract: LLM-based LEGO assembly requires both semantic grounding and physical feasibility. In this paper, we identify a data-induced failure mode, physhack, in which generated assemblies satisfy physical-validity constraints while remaining geometrically m...

📖 Read original article


368. Asymptotic Optimality of Thompson Sampling for Risk-Averse Bandits with Sub-Gaussian Rewards ​

Author: Joel Q. L. Chang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2606.09191v2 Announce Type: replace Abstract: We prove that $\rho\text{-}\mathrm{NPTS}_{\mathrm{SG}}$, an anchor-free nonparametric Thompson Sampling algorithm for risk-averse bandits, achieves regret matching the instance-dependent lower bound to leading order in $\log n$, establishing it as ...

📖 Read original article


369. SocraticPO: Policy Optimization via Interactive Guidance ​

Author: Zirui Liu, Tingyue Pan, Jie Ouyang, Qi Liu, Xianquan Wang, Jiayu Liu, Qingchuan Li, Jing Sha, Zhenya Huang, Shijin Wang, Enhong Chen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2606.09887v2 Announce Type: replace Abstract: Reinforcement learning (RL) for large language models usually supervises reasoning with scalar outcome rewards, such as binary correctness. Such rewards provide an optimization direction but rarely explain how a model should revise its mistaken rea...

📖 Read original article


370. TRACE: A Unified Rollout Budget Allocation Framework for Efficient Agentic Reinforcement Learning ​

Author: Heming Zou, Qi Wang, Yun Qu, Yuhang Jiang, Lizhou Cai, Yixiu Mao, Ru Peng, Xin Xu, Weijie Liu, Kai Yang, Saiyong Yang, Xiangyang Ji
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2606.11119v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) is a promising approach for enhancing reasoning and agentic behavior in large language models. However, rollout-intensive policy optimization is often limited by insufficient reward contrast, ar...

📖 Read original article


371. Chain of Operators: An Inference-Time Harness for In-Context Operator Learning ​

Author: Minghui Yang, Chenghan Wu, Ling Guo, Liu Yang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.12318v2 Announce Type: replace Abstract: While scientific foundation models show immense promise in accelerating physical simulations and numerical forecasting, they remain notoriously brittle when encountering out-of-distribution (OOD) scenarios. Adapting these generalist models to compl...

📖 Read original article


372. A Human-in-the-Loop Bayesian Optimization Framework for Constraint-Aware Bioprocess Development ​

Author: Samuel Stricker, Claus Wirnsperger, Alessandro Butt'e, Laura Helleckes, Gonzalo Guill'en Gos'albez, Antonio del Rio Chanona, Mehmet Mercang"oz
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, stat.ML

arXiv:2606.19230v2 Announce Type: replace Abstract: This work presents an extension to Pareto Front Guided Sampling (PFGS), a Human-in-the-Loop (HitL) Bayesian Optimization (BO) framework in which Gaussian process (GP) surrogate-derived quantities are reformulated as objectives of a multi-objective ...

📖 Read original article


373. What LLMs explain is not what they believe: Evaluating explanation sufficiency under models' own input beliefs ​

Author: Nhi Nguyen, Shauli Ravfogel, Rajesh Ranganath
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, stat.ML

arXiv:2606.28615v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in high-stakes domains, where free-text explanations such as chain-of-thought and post-hoc rationales are used to justify model outputs. Yet it remains unclear whether these explanations are su...

📖 Read original article


374. Group-Equivariant Poincar\'e Convolutional Networks ​

Author: Aiden Durrant, Rahul Baburajan, Georgios Leontidis
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.00556v2 Announce Type: replace Abstract: While recent methods like that of the Poincar'e ResNet have demonstrated the ability to learning visual representations directly in hyperbolic space, their optimisation remains a challenge, primarily due to the parameter redundancy of learning dis...

📖 Read original article


375. Post-Generation Curation of Synthetic Images via Homogeneous-Heterogeneous Splitting ​

Author: Disheng Liu, Tuo Liang, Chaoda Song, Yu Yin
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.02637v2 Announce Type: replace Abstract: Recent generative models can produce high-quality synthetic images, offering scalable training training data for data-hungry models. Existing approaches to exploiting this potential typically involve 1) training or fine-tuning generators, or 2) usi...

📖 Read original article


376. A JoLT for the KV cache: Near-lossless KV cache compression via joint Lagrangian allocation of Tucker ranks and a rotated residual for llms ​

Author: Rahul Krishnan, Volker Schulz
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, math.OC

arXiv:2607.12550v3 Announce Type: replace Abstract: The key-value (KV) cache has become the dominant memory cost of transformer inference: it grows with batch size, context length, and depth, and at long context it, rather than the model weights, sets the throughput ceiling. Existing reductions fall...

📖 Read original article


377. Long-term User Engagement Optimization through Model-agnostic Downstream Rewards Learning ​

Author: Dingsu Wang, Filip Ryzner, Kelly He, Armando Ordorica, David Woo, Aditya Mantha, Liyao Lu, Usha Amrutha Nookala, Haoran Guo, Jiacong He, Olafur Gudmundsson, Matt Chun, Krystal Benitez, Haibin Xie, Alekhya Pyla, Sameer Jain, Zhongjian Jiang, Shruthi Hariharan, Dhruvil Deven Badani, Yijie Dylan Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2607.14192v3 Announce Type: replace Abstract: As recommender systems mature in the past few years, their optimization objectives have evolved from a primary focusing on short-term behavioral signals to a broader emphasis on long-term user engagement and retention. However, directly optimizing ...

📖 Read original article


378. CardioMeta: Calibrated Multi-Task Prediction of Diabetes, Hypertension, and Cardiovascular Disease Across Population and EHR Data ​

Author: S M Asif Hossain, Ruksat Khan Shayoni, M. F. Mridha, Jungpil Shin
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.15721v2 Announce Type: replace Abstract: Cardiometabolic diseases remain among the most persistent drivers of preventable morbidity because diabetes, hypertension, and cardiovascular disease frequently co-occur and share metabolic, vascular, demographic, and behavioral determinants. Exist...

📖 Read original article


379. CANDOR: Chance-Calibrated Discordance in Frozen Foundation Encoders ​

Author: Soroosh Tayebi Arasteh, Sven Nebelung, Daniel Truhn
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CV

arXiv:2607.18451v2 Announce Type: replace Abstract: Frozen encoders are chosen by how well a lightweight head reads a finding from their features, not whether the geometry separates it. Nearest-neighbor discordance does, but with unequal banks the opposite-label neighbor wins on density, not geometr...

📖 Read original article


380. Do Sheaf Neural Networks Use Holonomy? A Measure--Intervene--Control Study ​

Author: Ankit Grover, R'emi Bourgerie
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19514v2 Announce Type: replace Abstract: Geometric architectures are often motivated by internal mechanisms, but accuracy alone does not show whether predictions use them. In Sheaf Neural Networks (SNNs), edge transports form a connection whose cycle products define holonomy. We ask wheth...

📖 Read original article


381. An Empirical Study of Feature Selection Granularity ​

Author: Muhammad Rajabinasab, Arthur Zimek
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.24145v2 Announce Type: replace Abstract: Feature selection aims to identify the most informative and relevant features for a given dataset, either in terms of capturing the underlying data structure and distribution better, or with respect to the performance on a downstream task. Existing...

📖 Read original article


382. A2TTA: Anchored-and-Agile Test-Time Adaptation for Evolving Traffic Sensor Networks ​

Author: Du Yin, Xiachong Lin, Yue Tan, Jinliang Deng, Estrid He, Hao Xue, Flora D. Salim
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25875v3 Announce Type: replace Abstract: Traffic forecasting is important for efficient traffic management and route planning in smart cities. Existing traffic forecasting studies typically assume fixed sensor graphs, overlooking the continuous evolution of real-world traffic networks, e....

📖 Read original article


383. Measuring in-context algorithmic reasoning in language models against an exact Bayes-optimal reference ​

Author: Luan Ozelim, Hector Zenil
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.01575v2 Announce Type: replace Abstract: Whether large language models perform algorithmic inference or pattern completion is hard to test, because most benchmarks supply answers but no distributional reference for what the shown evidence licenses. F-ICL supplies one exactly: we exhaustiv...

📖 Read original article


384. Control Allocation in Neural Network Optimization: Joint Affine Control of Weight and Bias Updates ​

Author: Zhang Gongyue, Sheng Yixuan, Wang Zhiyong, Liu Donghan, Ren Weihong, Liu Honghai
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02991v2 Announce Type: replace Abstract: Optimization algorithms determine not only the magnitude of a neural-network update but also how that update is distributed across parameter channels. We study whether this distribution can be treated as a controllable quantity independently of glo...

📖 Read original article


385. Provably Learning Multi-Head Attention with Queries ​

Author: Sunyeop Kim, Insung Kim
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.03294v3 Announce Type: replace Abstract: We study the problem of learning multi-head softmax attention from black-box input-output access. The learner may query arbitrary real-valued token sequences and observe only the scalar output at the final token. Recent work gives an algorithm usin...

📖 Read original article


386. When Calibration Depends on the Scoring Rule: Quantized Biomedical LLM Classification ​

Author: Anton Rasmussen, Hong Qin
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03854v3 Announce Type: replace Abstract: Quantized large language models enable on-premises processing of sensitive data, but their confidence estimates must be trustworthy. Reliability depends on implementation choices--prompt template, label wording, and scoring normalization--that are ...

📖 Read original article


387. SVI-DAG: A Structured Variational Inference Approach to Bayesian Causal Discovery ​

Author: Shrenik Zinage
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04930v2 Announce Type: replace Abstract: Bayesian causal discovery seeks to determine the posterior distribution of causal theories, which are interpreted as directed acyclic graphs (DAGs) that explain the observed data. The resulting posterior allows systematic reasoning regarding episte...

📖 Read original article


388. Edge Sparsification via Temporal Forman-Ricci Curvature for Dynamic Graph Learning ​

Author: Poupak Azad, Cuneyt Gurcan Akcora, Kiarash Shamsi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07158v2 Announce Type: replace Abstract: Temporal graph learning has become essential for analyzing real-world systems whose interactions continuously evolve over time, including financial transaction networks, communication systems, and online social platforms. However, learning from lar...

📖 Read original article


389. Reproducible Evaluation of MoE Expert Caching: Replay Semantics, Workload Contamination, and Operating Regimes ​

Author: Yu Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.PF

arXiv:2608.07911v3 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache management an attractive lever: a policy that raised the hit rate would cut expert traffic per t...

📖 Read original article


390. Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA ​

Author: Mind Lab, :, Vin Bo, Asher Cai, Jingwei Cao, Song Cao, Vic Cao, Amelia Chen, Andrew Chen, Kaijie Chen, Cleon Cheng, Steven Chiang, Kaixuan Fan, Hera Feng, Huan Feng, Arthur Fu, Aaron Guan, Jun Gao, Pyke Han, Nolan Ho, Ori Hong, Hailee Hou, Piers Hua, Charles Huang, Miles Jiang, Nora Jiang, Yuyi Jiang, Qiuyu Jin, Fancy Kong, Kuss Koo, Echo Lee, Jaron Lee, Andrew Lei, Alexy Li, Dawn Li, Lucian Li, Ray Li, Ricardo Li, Smith Li, Theo Li, Allen Lin, Elliot Lin, Fan Lin, Chen Ling, Kairus Liu, Kieran Liu, Logan Liu, Neo Liu, Xiang Liu, Yuxin Lu, Maeve Luo, Pony Ma, Verity Niu, Cole Qiao, Guian Qiu, Vince Qu, Sentry, Zhuoran Shen, Niko Song, Vincent Wang, Bo Wu, Rio Yang, Schacter Yang, Evelyn Ye, Fiona Ye, Ina Ye, Regis Ye, Josh Ying, Atlas Zeng, Danney Zeng, Salmon Zhan, Anya Zhang, Di Zhang, Mia Zhang, Sueky Zhang, Xuening Zhang, Wei Zhao, Ada Zhou, Adrian Zhou, Yuhua Zhou, Juno Zhu, Murphy Zhuang, Mindverse Team
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.09819v2 Announce Type: replace Abstract: Macaron-V1 is an open agent-model family for experiential intelligence: learning from experience in real environments and continuing to learn after deployment. It is organized around two system goals. Adaptation is pursued through recursive improve...

📖 Read original article


391. Reading the Gate, Not the Interference: Output-Side Interference Measurement Does Not Track Merge Collapse ​

Author: Chencheng Zhu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11797v2 Announce Type: replace Abstract: Model merging by task arithmetic works until it doesn't, and the field diagnoses why by measuring interference inside the merged model. We take the most direct such measure - the exact layerwise activation cross-term of a factorial ledger - establi...

📖 Read original article


392. Towards Understanding On-Policy Distillation through the Lens of Test-Time Scaling ​

Author: Xinmu Ge, Zizhuo Zhang, Yu Huang, Jianing Zhu, Lin Yuan, Wanli Gu, Weichang Wu, Weiran Huang, Xiaolu Zhang, Bo Han, Jun Zhou, Jiangchao Yao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.11829v2 Announce Type: replace Abstract: On-policy distillation (OPD) has emerged as a promising post-training technique for enhancing LLM reasoning. It is commonly believed to enable the student model to distill knowledge from a stronger teacher model, thereby expanding capabilities beyo...

📖 Read original article


393. Towards Truly Unsupervised Evaluation of Feature Selection ​

Author: Hafiz Saud Arshad, Muhammad Rajabinasab, Arthur Zimek
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12057v3 Announce Type: replace Abstract: Feature selection is one of the most important and fundamental tasks in data mining, tackled by a family of methods with an established set of evaluation techniques to measure the quality of a specific method. Most of the methods commonly used for ...

📖 Read original article


394. HI-MeshGraphNets: Efficient and Accurate Mesh-based Physics Learning with Hierarchical Multi-scale Graph Neural Networks ​

Author: SiHun Lee, Dong-Hyuk Park, Taesoo Bang, Seung-Hoon Kang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13827v2 Announce Type: replace Abstract: Machine-learned physical surrogate models have become promising alternatives to mesh-based numerical solvers. Among them, graph neural networks (GNNs) are well suited for representing simulation meshes and learning nodal state evolution through mes...

📖 Read original article


395. Rethinking Reverse KL as Adaptive Entropy Distillation ​

Author: Shizhen Li, Zhiyu Shen, Yuyin Lu, Yunhe Pang, Jielin Song, Yanghui Rao, Fu Lee Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.14685v2 Announce Type: replace Abstract: Knowledge distillation (KD) is widely used to transfer the capabilities of large language models (LLMs) to smaller students, but existing objectives often struggle to balance faithful imitation and robust generation. In particular, existing methods...

📖 Read original article


396. Structuring Semantic Embeddings for Principle Evaluation: A Prototype-Guided Contrastive Learning Approach ​

Author: Che Shen, Junwei Su, Lingpeng Kong, Chuan Wu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15224v2 Announce Type: replace Abstract: Reliable post-hoc evaluation asks whether already generated text satisfies a target criterion after generation. In this paper we study a focused frozen-embedding setting using principle-evaluation proxy tasks: toxicity detection, fine-grained emoti...

📖 Read original article


397. Towards a theory of inference-time alignment with unknown rewards ​

Author: Steve Hanneke, Hongao Wang, Mingyue Xu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15402v2 Announce Type: replace Abstract: Generative model alignment has received broad interest, and significant progress has been made in supervised fine-tuning and inference-time computation. Yet, alignment has remained poorly understood from a statistical learning perspective. We formu...

📖 Read original article


398. GraniKV: Asymmetric Granularity KV-Cache Paging for Multi-Agent Systems with Long Shared Prefix ​

Author: Jinhyun Jeon, Sungjoo Yoo
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15584v2 Announce Type: replace Abstract: Production paged-serving engines apply uniform paging granularity to the KV cache, even though the two regions of a multi-agent workload have opposite storage requirements: a long shared prefix demands contiguity, while the per-request suffix deman...

📖 Read original article


399. PertMind: Eliciting Emergent Biological Reasoning in LLM via Reinforcement Learning on Cellular Perturbation Data ​

Author: Zhenchao Tang, Xiaogang Xu, Tianxu Lv, Jiahui Guan, Jiale Zhou, Haohuai He, Zhi Song, Hanbo Huang, Jiehui Huang, Jiafei Wu, Zhe Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.QM

arXiv:2608.16419v2 Announce Type: replace Abstract: Large language models can describe mechanisms, yet scalable post-training still depends on costly, manually curated biological reasoning traces. Here we show that cellular perturbation atlases can instead become reinforcement-learning environments,...

📖 Read original article


400. Reference-free logged energy-oracle recovery for neural approximations of symmetric coercive variational problems: conforming Riesz reconstruction and archive-level selection ​

Author: Karim Bounja, Lahcen Laayouni, Boujemaa Achchab, Abdeljalil Sakat
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2608.16473v2 Announce Type: replace Abstract: Neural PDE training yields a finite checkpoint archive, yet its logged energy errors are inaccessible without the exact solution, while loss-based selection does not necessarily recover the logged energy oracle. For admissible neural approximations...

📖 Read original article


401. Learning to Unlearn: Machine Unlearning via Learning the Unlearning Behaviors ​

Author: Hang Zhang, Kaifeng Zhang, Yixiao Ma, Weijie Xu, Ye Zhu, Kai Ming Ting
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16700v2 Announce Type: replace Abstract: Various machine unlearning techniques have been developed in response to privacy legislation requirements, enabling individuals to exercise their legal right to have their data $D_f$ removed from a machine learning model. This process is typically ...

📖 Read original article


402. Elimination Geometry ​

Author: Mian Huang, Xueqin Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17646v2 Announce Type: replace Abstract: This monograph develops elimination geometry (EG), a typed, native-loss, audit-oriented framework for studying when locally optimal objects can be realized by a shared deployment rule. Elimination and compression may erase distinctions required by ...

📖 Read original article


403. Continuous Adversarial MeanFlow Transfer ​

Author: Yara Bahram, Zahra Dehghani, M'elodie Desbos, Eric Granger, Pablo Piantanida, Mohammadhadi Shateri
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.19540v2 Announce Type: replace Abstract: Training fast generators on new domains with limited data remains challenging for two reasons. First, adapting a pretrained diffusion or flow model to a new domain leaves its costly multi-step sampling unaddressed, and existing acceleration methods...

📖 Read original article


404. MileGPO: Milestone Inference with Local Evidence for Graph-Based Policy Optimization of Long-Horizon LLM Agents ​

Author: Bo Qian, Yuting Wu, Shuang Zeng, Huaiyu Wan, Dalin Zhang, Jiqiang Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.19803v2 Announce Type: replace Abstract: Credit assignment is challenging in long-horizon agentic reinforcement learning, where supervision often comes only from final rewards. Existing methods refine trajectory-level signals into step-level credits through step grouping or graph-based ad...

📖 Read original article


405. In-Cell Learning: Deployed Language Models Can Learn New Knowledge Without Changing a Single Stored Bit ​

Author: Zifeng Liu, Yaxin Lu, Xuanhan Wu, Zhiyong Du, Yiming Mao, Zhenhe Wang, Wenqi Shi, Zhengkun Jing, Linwei Liu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.20873v2 Announce Type: replace Abstract: A deployed language model is a file that many things depend on - a benchmark report, a certification, a fleet of devices - and every way of teaching it something new produces a different file. We show that this is not necessary. A 4-bit release sto...

📖 Read original article


406. BackDFL: A Unified Benchmark For Backdoor Attacks and Defenses In Decentralized Federated Learning ​

Author: Mouhamed Amine Bouchiha, Gregory Blanc, Yufei Han
Published: 8/25/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.DC

arXiv:2608.21137v2 Announce Type: replace Abstract: Decentralized Federated Learning (DFL) promises trust-free collaborative learning by replacing the centralized parameter server with peer-to-peer model exchange. However, this architectural shift fundamentally reshapes the threat landscape. Without...

📖 Read original article


407. Machine Learning Classification and Portfolio Construction: Does the Loss Function Matter? ​

Author: Yang Bai, Kuntara Pukthuanthong
Published: 8/25/2026, 4:00:00 AM
Categories: q-fin.GN, cs.LG, econ.GN, q-fin.CP, q-fin.EC, q-fin.PM

arXiv:2108.02283v4 Announce Type: replace-cross Abstract: Classification outperforms regression across matched machine learning models in portfolio construction. A stacking ensemble of gradient boosted tree, random forest, and neural network yields a value-weighted annualized Sharpe ratio of 1.83 fo...

📖 Read original article


408. Neuro-Causal Factor Analysis ​

Author: Alex Markham, Mingyu Liu, Bryon Aragam, Liam Solus
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2305.19802v2 Announce Type: replace-cross Abstract: Factor analysis (FA) is a statistical method for explaining how mutually dependent observed variables can be represented in terms of mutually independent latent factors, and it is widely used in the psychological, biological, and physical sci...

📖 Read original article


409. Deep Clustering Evaluation: How to Validate Internal Clustering Validation Measures ​

Author: Zeya Wang, Chenglong Ye
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2403.14830v2 Announce Type: replace-cross Abstract: Deep clustering partitions complex high-dimensional data using deep neural networks for clustering. It involves projecting data into lower-dimensional embeddings before partitioning, which embarks unique evaluation challenges. Traditional clu...

📖 Read original article


410. Robust performance metrics for imbalanced classification problems ​

Author: Hajo Holzmann, Bernhard Klar
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2404.07661v2 Announce Type: replace-cross Abstract: We show that established performance metrics in binary classification, such as Matthews' correlation coefficient (MCC), Cohen's $\kappa$, the F-score or the Jaccard similarity coefficient are not robust to class imbalance in the sense that if...

📖 Read original article


411. Memory-Enhanced Neural Solvers for Routing Problems ​

Author: Felix Chalumeau, Refiloe Shabe, Noah De Nicola, Arnu Pretorius, Thomas D. Barrett, Nathan Grinsztajn
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2406.16424v4 Announce Type: replace-cross Abstract: Routing Problems are central to many real-world applications, yet remain challenging due to their (NP-)hard nature. Amongst existing approaches, heuristics often offer the best trade-off between quality and scalability, making them suitable f...

📖 Read original article


412. Cross-validating causal discovery via Leave-One-Variable-Out ​

Author: Daniela Schkoda, Philipp Faller, Patrick Bl"obaum, Dominik Janzing
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2411.05625v2 Announce Type: replace-cross Abstract: We propose a new approach to falsify causal discovery algorithms without ground truth, which is based on testing the causal model on a variable pair excluded during learning the causal model. Specifically, given data on $X, Y, \boldsymbol{Z}=...

📖 Read original article


413. Conditional regression for the Nonlinear Single-Variable Model ​

Author: Yantao Wu, Mauro Maggioni
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2411.09686v4 Announce Type: replace-cross Abstract: Regressing a function $F$ on $\mathbb{R}^d$ without incurring the statistical and computational curse of dimensionality requires exploitable structure. Compositional models $F=f\circ g$ in which $g$ has a low-dimensional range include classic...

📖 Read original article


414. Semiparametric Double Reinforcement Learning with Applications to Long-Term Causal Inference ​

Author: Lars van der Laan, David Hubbard, Allen Tran, Nathan Kallus, Aur'{e}lien Bibaut
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2501.06926v5 Announce Type: replace-cross Abstract: Double reinforcement learning (DRL) provides efficient off-policy inference for policy values in nonparametric Markov decision processes (MDPs), but fully nonparametric estimators can be unstable when intertemporal overlap is weak and occupan...

📖 Read original article


415. AdaSemSeg: An Adaptive Few-shot Semantic Segmentation of Seismic Facies ​

Author: Surojit Saha, Ross Whitaker
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2501.16760v2 Announce Type: replace-cross Abstract: Automated interpretation of seismic images using deep learning methods is challenging because of the limited availability of training data. Few-shot learning is a suitable learning paradigm in such scenarios due to its ability to adapt to a n...

📖 Read original article


416. Cascaded Learned Bloom Filter for Optimizing Model-Filter Size Balance and Fast Rejection ​

Author: Atsuki Sato, Yusuke Matsui
Published: 8/25/2026, 4:00:00 AM
Categories: cs.DS, cs.CC, cs.LG

arXiv:2502.03696v2 Announce Type: replace-cross Abstract: Recent studies have demonstrated that learned Bloom filters (LBFs), which combine machine learning with the classical Bloom filter, can achieve superior memory efficiency. However, two challenges remain: (1) jointly optimizing the sizes of th...

📖 Read original article


417. Bringing Generative Learning to Representation Learning: Self-Supervised Transfer Learning as Distribution Matching ​

Author: Yuling Jiao, Wensen Ma, Defeng Sun, Hansheng Wang, Yang Wang
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, stat.ME

arXiv:2502.14424v4 Announce Type: replace-cross Abstract: Most self-supervised learning objectives defend against collapse but leave the target representation law unspecified. We formulate representation learning as Distribution Matching (DM), learning an augmentation-invariant encoder whose induced...

📖 Read original article


418. AI University: An LLM-Powered Learning Assistant for Engineering---A Finite Element Method Case Study ​

Author: Mostafa Faghih Shojaei, Rahul Gulati, Benjamin A. Jasperson, Shangshang Wang, Simone Cimolato, Manas Vardhan, Dangli Cao, Willie Neiswanger, Krishna Garikipati
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CL, cs.LG

arXiv:2504.08846v2 Announce Type: replace-cross Abstract: We introduce AI University (AI-U), a flexible framework for AI-driven course content delivery that adapts to a course's instructional style. AI-U combines a fine-tuned large language model (LLM) with retrieval-augmented generation (RAG) and a...

📖 Read original article


419. A Modular Multitask Reasoning Framework Integrating Spatio-temporal Models and LLMs ​

Author: Kethmi Hirushini Hettige, Jiahao Ji, Cheng Long, Shili Xiang, Gao Cong, Jingyuan Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2506.20073v2 Announce Type: replace-cross Abstract: Spatio-temporal data mining plays a pivotal role in informed decision making across diverse domains. However, existing models are often restricted to narrow tasks, lacking the capacity for multi-task inference and complex long-form reasoning ...

📖 Read original article


420. MixLoRA-DSI: Dynamically Expandable Mixture-of-LoRA Experts for Rehearsal-Free Generative Retrieval over Dynamic Corpora ​

Author: Tuan-Luc Huynh, Thuy-Trang Vu, Weiqing Wang, Trung Le, Dragan Ga\v{s}evi'c, Yuan-Fang Li, Thanh-Toan Do
Published: 8/25/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG

arXiv:2507.09924v2 Announce Type: replace-cross Abstract: Continually updating model-based indexes in generative retrieval with new documents remains challenging, as full retraining is computationally expensive and impractical under resource constraints. We propose MixLoRA-DSI, a novel framework tha...

📖 Read original article


421. Text-ADBench: Text Anomaly Detection Benchmark Based on LLM Embeddings ​

Author: Feng Xiao, Jicong Fan
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2507.12295v2 Announce Type: replace-cross Abstract: Text anomaly detection is a critical task in natural language processing (NLP), with applications spanning fraud detection, misinformation identification, spam detection and content moderation, etc. Despite significant advances in large langu...

📖 Read original article


422. Higher-Order Kuramoto Oscillator Network for Dense Associative Memory ​

Author: Jona Nagerl, Natalia G. Berloff
Published: 8/25/2026, 4:00:00 AM
Categories: nlin.AO, cond-mat.dis-nn, cond-mat.stat-mech, cs.ET, cs.LG

arXiv:2507.21984v2 Announce Type: replace-cross Abstract: Networks of phase oscillators can serve as dense associative memories when they incorporate genuine many-body coupling beyond the classical Kuramoto model's pairwise interaction. Here we introduce a generalized Hebbian Kuramoto model that com...

📖 Read original article


423. Model Directions, Not Words: Mechanistic Topic Models Using Sparse Autoencoders ​

Author: Carolina Zheng, Nicolas Beltran-Velez, Sweta Karlekar, Claudia Shi, Achille Nazaret, Asif Mallik, Amir Feder, David M. Blei
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2507.23220v3 Announce Type: replace-cross Abstract: Traditional topic models are effective at uncovering latent themes in large text collections. However, due to their reliance on bag-of-words representations, they struggle to capture semantically abstract features. While some neural variants ...

📖 Read original article


424. Models in the Same Family are NOT Trust-Equivalent ​

Author: Rohit Raj Rai, Chirag Kothari, Siddhesh Shelke, Yatika Jena, Amit Awekar
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2508.13533v2 Announce Type: replace-cross Abstract: Within a model family, a smaller variant is often deployed as a drop-in replacement for a larger one when their performance is similar. However, performance alone does not tell the full story. We propose a framework to evaluate trust-equivale...

📖 Read original article


425. Programmable k-local Ising interactions and shallow optical Kolmogorov--Arnold networks through repeated data encounters ​

Author: Nikita Stroev, Natalia G. Berloff
Published: 8/25/2026, 4:00:00 AM
Categories: physics.optics, cs.ET, cs.LG

arXiv:2508.17440v3 Announce Type: replace-cross Abstract: Photonic processors are naturally suited to linear transformations, but independently programmable higher-order interactions usually require nonlinear media or a reduction to pairwise models. We introduce a repeated-encounter architecture tha...

📖 Read original article


426. Benchmarking Retrieval-Augmented Generation Strategies for Large Language Model-Based Travel Mode Choice Prediction ​

Author: Yiming Xu, Junfeng Jiao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG

arXiv:2508.17527v2 Announce Type: replace-cross Abstract: Accurately predicting travel mode choice is essential for effective transportation planning, yet traditional statistical and machine learning models are constrained by rigid assumptions, limited contextual reasoning, and reduced transferabili...

📖 Read original article


427. Optimal Transport for Handwritten Text Recognition in a Low-Resource Regime ​

Author: Petros Georgoulas Wraight, Giorgos Sfikas, Ioannis Kordonis, Petros Maragos, George Retsinas
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2509.16977v2 Announce Type: replace-cross Abstract: Handwritten Text Recognition (HTR) is a task of central importance in the field of document image understanding. State-of-the-art methods for HTR require the use of extensive annotated sets for training, making them impractical for low-resour...

📖 Read original article


428. Minimum Bisection Problem: Machine Learning-Based Penalty Parameter Tuning for Optimization on Quantum Annealers ​

Author: Ren'ata Rusn'akov'a, Martin Chovanec, Juraj Gazda
Published: 8/25/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2509.19005v2 Announce Type: replace-cross Abstract: The Minimum Bisection Problem is a fundamental, computationally hard graph partitioning problem with applications in parallel computing, network design, and large-scale data processing. When formulated as a Quadratic Unconstrained Binary Opti...

📖 Read original article


429. Training Proactive and Personalized LLM Agents ​

Author: Weiwei Sun, Xuhui Zhou, Weihua Du, Xingyao Wang, Sean Welleck, Graham Neubig, Maarten Sap, Yiming Yang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2511.02208v2 Announce Type: replace-cross Abstract: Despite rapid progress, current AI agents are primarily optimized for isolated task completion. We argue for a paradigm shift toward training agents as collaborators that communicate and adapt to people. To facilitate this shift in real-world...

📖 Read original article


430. DIGing--SGLD: Decentralized and Scalable Langevin Sampling over Time--Varying Networks ​

Author: Waheed U. Bajwa, Mert Gurbuzbalaban, Mustafa Ali Kutbay, Lingjiong Zhu, Muhammad Zulqarnain
Published: 8/25/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ML

arXiv:2511.12836v2 Announce Type: replace-cross Abstract: Sampling from a target distribution induced by training data is central to Bayesian learning, with Stochastic Gradient Langevin Dynamics (SGLD) serving as a key tool for scalable posterior sampling and decentralized variants enabling learning...

📖 Read original article


431. ADHint: Adaptive Hints with Difficulty Priors for Reinforcement Learning ​

Author: Feng Zhang, Zezhong Tan, Xinhong Ma, Ziqiang Dong, Xi Leng, Jianfei Zhao, Xin Sun, Yang Yang, Guanjun Jiang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2512.13095v3 Announce Type: replace-cross Abstract: To address the limited capability expansion and low sample efficiency of Reinforcement Learning (RL), recent methods have integrated ''hints'' into post-training, which are prefix segments of complete reasoning trajectories, aiming for powerf...

📖 Read original article


432. Hierarchical Book Organization for Learning-Resource Discovery using Dual-Path Graph Convolutions ​

Author: Suraj Kumar, Utsav Kumar Nareti, Soumi Chattopadhyay, Chandranath Adak, Prolay Mallick
Published: 8/25/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, cs.MM

arXiv:2512.21076v2 Announce Type: replace-cross Abstract: The growing availability of books and textual materials in digital learning environments necessitates reliable semantic organization to support scalable resource management and discovery. However, existing book classification approaches typic...

📖 Read original article


433. An Empirical Study on Preference Tuning Generalization and Diversity Under Domain Shift ​

Author: Constantinos Karouzos, Xingwei Tan, Nikolaos Aletras
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2601.05882v2 Announce Type: replace-cross Abstract: Preference tuning aligns base language models to human judgments of quality, helpfulness, or safety by optimizing over explicit preference signals rather than likelihood alone. Prior work has shown that preference tuning degrades performance ...

📖 Read original article


434. LLM-Based Adversarial Persuasion Attacks on Fact-Checking Systems ​

Author: Jo~ao A. Leite, Olesya Razuvayevskaya, Kalina Bontcheva, Carolina Scarton
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2601.16890v2 Announce Type: replace-cross Abstract: Automated fact-checking (AFC) systems are susceptible to adversarial attacks, enabling false claims to evade detection. Existing adversarial frameworks typically rely on injecting noise or altering semantics, yet no existing framework exploit...

📖 Read original article


435. Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints ​

Author: Ruoyu Chen, Shangquan Sun, Xiaoqing Guo, Kangwei Liu, Sanyi Zhang, Zhangcheng Wang, Shiming Liu, Qunli Zhang, Wei Wang, Hua Zhang, Xiaochun Cao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2602.07008v5 Announce Type: replace-cross Abstract: Reliable models should not only predict correctly, but also base their decisions on acceptable evidence. However, conventional supervised learning typically provides only class-level labels, allowing models to achieve high accuracy by exploit...

📖 Read original article


436. Mode Seeking meets Mean Seeking for Fast Long Video Generation ​

Author: Shengqu Cai, Weili Nie, Chao Liu, Julius Berner, Lvmin Zhang, Nanye Ma, Hansheng Chen, Maneesh Agrawala, Leonidas Guibas, Gordon Wetzstein, Arash Vahdat
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2602.24289v2 Announce Type: replace-cross Abstract: Scaling video generation from seconds to minutes faces a critical bottleneck: while short-video data is abundant and high-fidelity, coherent long-form data is scarce and limited to narrow domains. To address this, we propose a training paradi...

📖 Read original article


437. SkillNet: Create, Evaluate, and Connect AI Skills ​

Author: Yuan Liang, Ruobin Zhong, Haoming Xu, Chen Jiang, Yi Zhong, Runnan Fang, Jia-Chen Gu, Shumin Deng, Yunzhi Yao, Mengru Wang, Shuofei Qiao, Yida Xue, Xin Xu, Tongtong Wu, Kun Wang, Yang Liu, Zhen Bi, Jungang Lou, Yuchen Eleanor Jiang, Hangcheng Zhu, Gang Yu, Haiwen Hong, Longtao Huang, Hui Xue, Chenxi Wang, Yijun Wang, Zifei Shan, Xi Chen, Zhaopeng Tu, Feiyu Xiong, Xin Xie, Peng Zhang, Zhengke Gui, Lei Liang, Jun Zhou, Chiyu Wu, Jin Shang, Yu Gong, Junyu Lin, Changliang Xu, Hongjie Deng, Wen Zhang, Keyan Ding, Qiang Zhang, Fei Huang, Ningyu Zhang, Jeff Z. Pan, Guilin Qi, Haofen Wang, Huajun Chen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CV, cs.LG, cs.MA

arXiv:2603.04448v3 Announce Type: replace-cross Abstract: Current AI agents can flexibly invoke tools and execute complex tasks, yet their long-term advancement is hindered by the lack of systematic accumulation and transfer of skills. Without a unified mechanism for skill consolidation, agents freq...

📖 Read original article


438. Agentic-Kube: A Graph-Enhanced Multi-Agent Reinforcement Learning Framework for Multi-Objective Kubernetes Scheduling ​

Author: Hamed Hamzeh
Published: 8/25/2026, 4:00:00 AM
Categories: cs.DC, cs.LG, cs.MA

arXiv:2603.12031v3 Announce Type: replace-cross Abstract: Cloud-native container orchestration requires resource schedulers capable of balancing infrastructure expenditure, fault resilience, and node utilisation. Conventional reinforcement learning approaches typically rely on monolithic single-agen...

📖 Read original article


439. ZOTTA: Test-Time Adaptation with Gradient-Free Zeroth-Order Optimization ​

Author: Ronghao Zhang, Shuaicheng Niu, Qi Deng, Yanjie Dong, Jian Chen, Runhao Zeng
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2603.14254v2 Announce Type: replace-cross Abstract: Test-time adaptation (TTA) aims to improve model robustness under distribution shifts by adapting to unlabeled test data, but most existing methods rely on backpropagation (BP), which is computationally costly and incompatible with non-differ...

📖 Read original article


440. When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making ​

Author: Jun Liu, Pu Zhao, Zhenglun Kong, Xuan Shen, Peiyan Dong, Fan Yang, Lin Cui, Hao Tang, Geng Yuan, Wei Niu, Wenbin Zhang, Xue Lin, Gaowen Liu, Yanzhi Wang, Dong Huang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2603.16673v5 Announce Type: replace-cross Abstract: Embodied robotic systems increasingly rely on large language model (LLM)-based agents to support high-level reasoning, planning, and decision-making during interactions with the environment. However, invoking LLM reasoning introduces substant...

📖 Read original article


441. UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference ​

Author: Lang Zhou, Shuxuan Li, Zhuohao Li, Shi Liu, Wei-Shi Zheng, Zhilin Zhao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2603.18446v3 Announce Type: replace-cross Abstract: Long-context inference remains challenging for large language models due to attention dilution and out-of-distribution degradation. Context selection mitigates this limitation by attending to a subset of key-value cache entries, yet most meth...

📖 Read original article


442. Robust Multi-Agent Reinforcement Learning for Small UAS Separation Assurance under GPS Degradation and Spoofing ​

Author: Alex Zongo, Filippos Fotiadis, Ufuk Topcu, Peng Wei
Published: 8/25/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, cs.SY, eess.SY

arXiv:2603.28900v2 Announce Type: replace-cross Abstract: We address robust separation assurance for small Unmanned Aircraft Systems (sUAS) under GPS degradation and spoofing via Multi-Agent Reinforcement Learning (MARL). In cooperative surveillance, each aircraft (or agent) broadcasts its GPS-deriv...

📖 Read original article


443. Continuous Orthogonal Mode Decomposition: Haptic Signal Prediction in Tactile Internet ​

Author: Mohammad Ali Vahedifar, Mojtaba Nazari, Qi Zhang
Published: 8/25/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2604.09446v2 Announce Type: replace-cross Abstract: The Tactile Internet demands sub-millisecond latency and ultra-high reliability, as even slight latency or packet loss can destabilize haptic control. To address this, we propose the Mode-Domain Architecture (MDA), a bilateral predictive neur...

📖 Read original article


444. Large Language Models Generate Harmful Responses Using a Distinct Mechanism, Shared Across Harm Types ​

Author: Hadas Orgad, Boyi Wei, Kaden Zheng, Martin Wattenberg, Peter Henderson, Seraphina Goldfarb-Tarrant, Yonatan Belinkov
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.09544v3 Announce Type: replace-cross Abstract: Large language models remain vulnerable to jailbreaks that elicit harmful responses, yet the mechanism behind harmful response generation is poorly understood. Here, we investigate how this capability is organized within model parameters. We ...

📖 Read original article


445. MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction ​

Author: Wenchang Duan, Zhenguo Gao, Jinguo Xian, Yi Shi
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2604.10169v4 Announce Type: replace-cross Abstract: Trajectory prediction is a key component of autonomous driving systems because future motions directly affect collision checking, behavior planning, and control. The task remains challenging under dense interactions, heterogeneous behaviors, ...

📖 Read original article


446. DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories ​

Author: Neemesh Yadav, Palakorn Achananuparp, Jing Jiang, Ee-Peng Lim
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.20443v3 Announce Type: replace-cross Abstract: We introduce DialToM, an annotated Theory of Mind (ToM) benchmark built from naturalistic human-human dialogues using a multiple-choice evaluation framework. Concurrent with recent work showing a gap between explicit mental-state inference an...

📖 Read original article


447. Revisiting the Effectiveness of LLM Pruning for Test-Time Scaling ​

Author: Ocean Monjur, Shahriar Kabir Nahin, Anshuman Chhabra
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2604.25098v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) now exhibit remarkable reasoning capabilities through test-time compute scaling (TTS), with impressive performance across math and coding benchmarks. In parallel, research in model compression has developed prunin...

📖 Read original article


448. Trident: Improving Malware Detection with LLMs and Behavioral Features ​

Author: Rebecca Saul, Jingzhi Jiang, Elliott Chia, David Wagner
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2605.00297v2 Announce Type: replace-cross Abstract: Traditionally, machine learning methods for PE malware detection have relied on static features like byte histograms, string information, and PE header contents. One barrier to incorporating dynamic analysis features has been the semi-structu...

📖 Read original article


449. FinSTaR: Towards Financial Reasoning with Time Series Reasoning Models ​

Author: Seunghan Lee, Jun Seo, Jaehoon Lee, Sungdong Yoo, Minjae Kim, Tae Yoon Lim, Dongwan Kang, Hwanil Choi, Soonyoung Lee, Wonbin Ahn
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2605.03460v5 Announce Type: replace-cross Abstract: Time series (TS) reasoning models (TSRMs) have shown promising capabilities in general domains, yet they consistently fail on financial domain, which exhibit unique characteristics. We propose a general 2 x 2 capability taxonomy for TSRMs by ...

📖 Read original article


450. When KV Meets Embeddings: Dynamic GPU Memory Allocation for Accelerating Generative Recommender Serving ​

Author: Wenjun Yu, Shuguang Han, Amelie Chi Zhou
Published: 8/25/2026, 4:00:00 AM
Categories: cs.DC, cs.IR, cs.LG

arXiv:2605.04450v2 Announce Type: replace-cross Abstract: Generative Recommender (GR) inference places embedding hot caches (EMB) and KV caches in direct competition for limited GPU HBM: allocating more memory to one improves its efficiency but degrades the other. Existing systems optimize them in i...

📖 Read original article


451. Stability of the Monge Map in Semi-Dual Optimal Transport ​

Author: Anton Selitskiy, David Millard
Published: 8/25/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2605.05569v4 Announce Type: replace-cross Abstract: This paper shows that the semi-dual formulation of the optimal transport problem has a degenerate saddle-point structure, and that its numerical solution is equivalent to solving a constrained optimization problem. We derive necessary and suf...

📖 Read original article


452. A Mixture Autoregressive Image Generative Model on Quadtree Regions for Gaussian Noise Removal via Variational Bayes and Gradient Methods ​

Author: Shota Saito, Yuta Nakahara, Kohei Horinouchi, Naoki Ichijo, Manabu Kobayashi, Toshiyasu Matsushima
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2605.11585v2 Announce Type: replace-cross Abstract: This paper addresses the problem of image denoising for grayscale images. We propose a probabilistic image generative model that combines a quadtree region-partitioning model with a mixture autoregressive model, and propose a framework that r...

📖 Read original article


453. GIM: Evaluating models via tasks that integrate multiple cognitive domains ​

Author: Rohit Patel, Alexandre Rezende, Steven McClain
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2605.18663v2 Announce Type: replace-cross Abstract: As LLM benchmarks saturate, the evaluation community has pursued two strategies to increase difficulty: escalating knowledge demands (GPQA, HLE) or removing knowledge entirely in favor of abstract reasoning (ARC-AGI). The first conflates memo...

📖 Read original article


454. MOSAIC: Modular Orchestration for Structured Agentic Intelligence and Composition ​

Author: Yifan Bao, Xinyu Xi, Xinyu Liu, Wen Ge, Lei Jiang, Kevin Zhang, Raad Khraishi, Yihao Ang, Anthony K. H. Tung, Lukasz Szpruch, Hao Ni
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.00708v2 Announce Type: replace-cross Abstract: Automated data science is a structured model-selection problem. A solution must choose data transformations, feature representations, architecture, training procedure, evaluation protocol, and refinement strategy for a task. AutoML systems au...

📖 Read original article


455. Practical and Optimal Algorithm for Linear Contextual Bandits with Rare Parameter Updates ​

Author: Sanghoon Yu, Min-hwan Oh
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2606.00984v3 Announce Type: replace-cross Abstract: We study linear contextual bandits under rare parameter updates: the learner may incorporate reward feedback into its parameter estimate only at a small number of update times, while still observing contexts online and selecting actions seque...

📖 Read original article


456. FlatVPR: Plug-and-play Geo-linear Residual Adapter for Geometric Rectification of Foundation Model Feature Manifolds ​

Author: Rai Hisada, Kanji Tanaka
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO

arXiv:2606.01734v2 Announce Type: replace-cross Abstract: This paper proposes ``FlatVPR,'' a novel geometric rectification paradigm that effectively bridges the trade-off between map lightweightness and localization accuracy in visual place recognition (VPR) by enforcing a feature manifold structure...

📖 Read original article


457. Automatic, Debiased, and Invariant Counterfactual Generation under General Interventions ​

Author: Raphael C Kim, Jingsen Zhu, Ramin Zabih, Michele Santacatterina
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2606.07399v2 Announce Type: replace-cross Abstract: Generative models for counterfactual outcomes have great potential to support decision-making under complex interventions, but existing approaches are limited by unstable estimation, poor generalization across environments, and bias from nuis...

📖 Read original article


458. How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions ​

Author: Donghao Huang, Tomas Drietomsky, Benjamin Barrett, Zhaoxia Wang
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.08051v2 Announce Type: replace-cross Abstract: Merchant information extraction turns noisy financial transaction descriptors into structured fields at production scale. Our deployed LoRA-fine-tuned LLaMA~3.1-8B reaches 96.95% F1, but its memory and throughput motivate smaller replacement...

📖 Read original article


459. Adjoint Method versus Physics-Informed Neural Networks in PDE-Constrained Inverse Problems ​

Author: Zhen Zhang, Alessandro Alla, George Em Karniadakis
Published: 8/25/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA

arXiv:2606.12337v2 Announce Type: replace-cross Abstract: Inverse problems governed by partial differential equations (PDEs) are central to computational mechanics and are commonly solved by adjoint-based optimization, while physics-informed neural networks (PINNs) have emerged as a flexible alterna...

📖 Read original article


460. Relational Structural Causal Models ​

Author: Adiba Ejaz, Elias Bareinboim
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SI, stat.ML

arXiv:2606.14892v2 Announce Type: replace-cross Abstract: An artificial intelligence must have a model of its environment that is causal, supporting reasoning about interventions and counterfactuals, and also combinatorial, supporting generalization to unseen combinations of objects. In this work, w...

📖 Read original article


461. Polynomial-Time Mistake-Bounded Language Generation ​

Author: H'ector Jimenez, Alexander Kozachinskiy, Vicente Opazo
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CC, cs.LG

arXiv:2606.16077v2 Announce Type: replace-cross Abstract: In this paper, we introduce a polynomial-time version of the mistake-bounded language generation (MBLG) framework due to Kleinberg, Peale, and Reingold (2026). We obtain upper and lower bounds for a number of simple families of Boolean functi...

📖 Read original article


462. Event-Conditioned Diagnostics of Kinematic, Contact, and Object-Permanence Structure in Passive Object-State World Models ​

Author: Yang Liu, Yuming Chen
Published: 8/25/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2606.28455v3 Announce Type: replace-cross Abstract: World models can predict future physical states, but prediction accuracy alone does not explain how physical information is organized and used inside their latent dynamics. We introduce a controlled diagnostic protocol for studying event-cond...

📖 Read original article


463. H-OPD: Confidence Aware Heterogeneous Multi-Teacher Multimodal On-policy Distillation ​

Author: Qixiang Yin, Huanjin Yao, Yuchen Cai, Jianghao Chen, Ziyi Wang, Min Yang, Fei Su, Zhicheng Zhao
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.02592v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) has recently emerged as an effective post-training paradigm by providing supervision on student-generated trajectories. However, existing OPD methods for multimodal reasoning usually rely on a static teacher routi...

📖 Read original article


464. Adversarial Robustness of Phishing Email Detection: A Comparative Study of TF-IDF + Logistic Regression and Fine-Tuned DistilBERT ​

Author: Tanveer Ahmed, Seyedali Pourmoafil
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CY, cs.LG

arXiv:2607.18429v2 Announce Type: replace-cross Abstract: Phishing emails remain one of the most persistent cybersecurity threats, and machine-learning classifiers are widely used to detect them. Most reported detection accuracies, however, are measured on clean, in-distribution test data rather tha...

📖 Read original article


465. CacheSpec: Finding the Sweet Spot for Small Models in Large Language Models ​

Author: Jingquan Chen, Jie Feng, Jinghua Piao, Shaogang Hu, Yong Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.20507v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for program-aided reasoning, agentic decision making, and structured task execution, but these settings often incur substantial inference cost. Many such requests share similar computational ...

📖 Read original article


466. CausalSmith: A Formally Grounded, Self-Improving Agentic Framework for Automated Research in Causal Inference ​

Author: Jiyuan Tan, Vasilis Syrgkanis
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, econ.EM

arXiv:2607.22511v3 Announce Type: replace-cross Abstract: Automating theoretical research is constrained not only by the generation of candidate results, but also by their reliable evaluation. A common approach is to close the research loop with a large language model (LLM) reviewer. However, such r...

📖 Read original article


467. On Non-Stationary Dynamic Pricing: Adaptivity and Optimality ​

Author: Feiyu Jiang, Zifeng Zhao
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.24115v2 Announce Type: replace-cross Abstract: We study the contextual dynamic pricing problem under non-stationarity, where a firm sells products to $T$ sequentially arriving consumers that behave according to an unknown demand model that can change over time. The demand model is assumed...

📖 Read original article


468. Few-Shot Open-Set Audio Classification via Transductive Prototype Refinement and Class Logit Enhancement ​

Author: Tianyan Deng, Yanxiong Li, Rui Gao, Jiahao Du
Published: 8/25/2026, 4:00:00 AM
Categories: cs.SD, cs.LG

arXiv:2607.26607v2 Announce Type: replace-cross Abstract: Few-shot Open-set audio classification requires classifying query samples from known classes with a few labeled support samples while rejecting query samples from unknown classes. Transductive inference jointly observes the full unlabeled que...

📖 Read original article


469. Nova: An End-to-End MLIR Compiler for Deep Learning ​

Author: Adwaid Suresh, Aparna A, Harshini V M, Jona Delcy C A, Killi Uma Maheswara Rao, Ram Charan Golla, Surendra Vendra
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.AR, cs.LG, cs.PL

arXiv:2608.00029v2 Announce Type: replace-cross Abstract: The performance of deep learning models at scale relies heavily on how effectively high-level mathematical operations are mapped to underlying physical hardware. While high-level tensor frameworks provide flexible abstractions, their executio...

📖 Read original article


470. Generative Models for Modeling and Synthesizing MIMO Channels in Adverse Weather Conditions ​

Author: Vignesh Nandakumar, Faraz Barati, Brian L. Evans
Published: 8/25/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.00156v2 Announce Type: replace-cross Abstract: The push for broader coverage in future cellular networks depends on reliable service, yet this is increasingly harder to do as we encounter more instances of extreme weather conditions. In extreme weather conditions, we have difficulty evalu...

📖 Read original article


471. Sparse PPMI Graph Averaging for Random Indexing Embeddings ​

Author: Sriram Loganathan, Gokul Anand, Aung Bo Bo, Yourui Shao, William B. Andreopoulos
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.05724v2 Announce Type: replace-cross Abstract: We study a specific sparse post-processing pipeline for Random Indexing (RI) on kinship analogies in a small fairytales corpus. The published artifacts use uniform RI context accumulation with 200 dimensions and eight nonzeros, followed by on...

📖 Read original article


472. Population-Scalable Multi-Agent World Modeling ​

Author: Renjie Zhao, Yuxiang Wu, Mingyu Zhang, Jiaxin Li, Sisi Li, He Li, Yimin Sheng, Tianxi Tan, Zhenkai Zhang, Jiao Liang, Jianyi Zhu, Yong-Lu Li
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.08600v3 Announce Type: replace-cross Abstract: World models have recently achieved impressive progress in visual prediction and interactive generation, but extending them to multi-agent environments introduces a fundamental scalability challenge. Existing methods generally assume a fixed ...

📖 Read original article


473. Do LLM Recommenders Know When They're Hallucinating? Auditing Confidence Calibration in Catalog Faithfulness ​

Author: Srijith Ravikumar
Published: 8/25/2026, 4:00:00 AM
Categories: cs.IR, cs.CL, cs.LG

arXiv:2608.10008v3 Announce Type: replace-cross Abstract: LLM recommenders for top-K item suggestion regularly emit titles outside the target catalog. Prior audits report a binary out-of-domain rate; none ask whether the model knew. We jointly audit hallucination rate (OOD@10) and verbalized-confide...

📖 Read original article


474. Evaluation Resolution Confounds Learning-Rule Comparisons in Model-Brain RSA of Early Visual Cortex ​

Author: Nils Leutenegger
Published: 8/25/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG

arXiv:2608.12408v2 Announce Type: replace-cross Abstract: Representational similarity analysis (RSA) is increasingly used to ask which learning rules give convolutional networks brain-like representations. Because biologically plausible rules such as feedback alignment, predictive coding and STDP do...

📖 Read original article


475. Maintaining IoT Device Identification under Concept Drift via Budget-Aware Traffic Labeling ​

Author: Shayan Azizi, Norihiro Okui, Masataka Nakahara, Ayumu Kubota, Gustavo Batista, Hassan Habibi Gharakaheili
Published: 8/25/2026, 4:00:00 AM
Categories: cs.NI, cs.CR, cs.LG

arXiv:2608.15465v2 Announce Type: replace-cross Abstract: Identification of IoT device types from passive traffic is increasingly used for security management in enterprise and ISP networks. However, the performance of machine learning-based classifiers gradually degrades under concept drift as devi...

📖 Read original article


476. Towards Zero-Shot Task Transfer with Neurosymbolic World Models ​

Author: Isidoro Tamassia, Lennert De Smet, Giuseppe Marra
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.17959v2 Announce Type: replace-cross Abstract: State-of-the-art model-based reinforcement learning methods learn neural world models that allow policy improvement by planning in a latent space, without assumptions on the structure of the underlying environment. While expressive, these mod...

📖 Read original article


477. GigaBrain-WBC-0.5: A Behavior World Model for Robust Whole-Body Control with Environment Interaction ​

Author: Ziyang Cheng, Tianshu Tang, Jinxin Lan, Xinze Chen, Yuhan Gong, Zhichao Liu, Changzhong Wu, Yahao Mao, Zongyan Deng, Mingxuan Ma, Huasen Xi, Yilong Liu, Yutong Wu, Xiaofeng Wang, Yang Wang, Yun Ye, Guan Huang, Xiaojie Jin, Zheng Zhu, Jiwen Lu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2608.18234v2 Announce Type: replace-cross Abstract: Whole-body motion tracking policies turn a humanoid into a robust control interface: the teleoperator---or an upstream model---only supplies a coarse movement intent, while the low-level policy keeps the robot balanced and physically feasible...

📖 Read original article


478. Sobolev Regularized Score Difference Estimation in Diffusion Models ​

Author: Chenghan Xie, Jose Blanchet, Renyuan Xu
Published: 8/25/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC

arXiv:2608.18237v2 Announce Type: replace-cross Abstract: Estimating the difference of two Stein's score functions is a fundamental problem in generative modeling. In particular, score differences arise naturally in transfer learning, where the score difference provides the mechanism for adapting a ...

📖 Read original article


479. Time-Series Retrieval for Grounding Multimodal Language Models in Remaining Useful Life Prediction ​

Author: Valeriu Dimidov, Rapha"el Frank
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.19218v2 Announce Type: replace-cross Abstract: Large language models (LLMs) and agentic AI systems are increasingly being explored for domain-specific maintenance and prognostics tasks, raising the question of whether they can effectively support prognostics and health management (PHM). I...

📖 Read original article


480. Active Spiking Perception: The Membrane Potential as a Belief State for Anytime 3D Point Cloud Recognition ​

Author: Akarsh Jain, Arya Pawa, Ayush Debnath, Smera Rawal, Sayeed Shafayet Chowdhury
Published: 8/25/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG

arXiv:2608.19232v2 Announce Type: replace-cross Abstract: Spiking point cloud networks usually scan space in a fixed, input-agnostic order, which leaves the most distinctive resource of spiking computation, the temporal evolution of the membrane potential, unused as a locus of decision-making. Activ...

📖 Read original article


481. Data-Driven Time-Varying Control Barrier Functions for Adaptive Safe-Set Learning with Online Decremental Support Vector Machines ​

Author: Shawon Dey, Michael Budihartono, Hever Moncayo
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CE, cs.LG

arXiv:2608.19366v2 Announce Type: replace-cross Abstract: Mission-critical intelligent systems often operate under time-varying limitations that reduce control authority and change the admissible safe operating envelope. In such settings, a safety certificate learned under nominal conditions may bec...

📖 Read original article


482. Question-Guided Evidence Acquisition for Multimodal Visual Question Answering ​

Author: Alin-Ionut Popa
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.19739v2 Announce Type: replace-cross Abstract: Multimodal LLMs can see a document, but they often can't read it reliably. Small text, tables, visual cues, and topological elements still trip them up under direct visual inference, even when the page is already sitting in the model's contex...

📖 Read original article


483. What You Can't See Is What You Learn: Restricted Evidence Visibility Favors Compositional Generalization in Shared-Genome Language-Model Societies ​

Author: Narcis Marincat
Published: 8/25/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2608.20054v2 Announce Type: replace-cross Abstract: Multi-module systems often expose every module to the full input. We test whether restricting evidence visibility changes which solutions gradient-based training discovers. Four-cell societies share one frozen pretrained language model and on...

📖 Read original article


484. Task-CoEvolve: Efficient Harness Optimization via Adaptive Validation Task Selection ​

Author: Atsuyuki Miyai, Kiyoharu Aizawa, Toshihiko Yamasaki
Published: 8/25/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.20169v2 Announce Type: replace-cross Abstract: We present a novel approach to efficient LLM harness optimization through adaptive validation task selection. Harness optimization iteratively rewrites the harness code based on validation performance, enabling substantial performance gains w...

📖 Read original article


485. Which Eviction Policy Should an LLM Cache Use? A Systematic Study Across Workloads, Capacities, and Encoders ​

Author: Yash Kulkarni, Shubham Harkare, Arvind Yogesh Suresh Babu
Published: 8/25/2026, 4:00:00 AM
Categories: cs.DB, cs.LG

arXiv:2608.20280v2 Announce Type: replace-cross Abstract: Semantic caches reuse an LLM response when the incoming query embedding lies near a cached query, but proposed eviction policies have rarely been compared under one protocol. Using CLEVER, we evaluate FIFO, LRU, LFU, ARC, GDSF, a single-pass ...

📖 Read original article