Skip to content

arXiv cs.LG - 2026-08-11 ​

570 items collected.


1. Application of Artificial Intelligence for Fraudulent Banking Operations Recognition ​

Author: Bohdan Mytnyk, Oleksandr Tkachyk, Nataliya Shakhovska, Solomiia Fedushko, Yuriy Syerov
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE, cs.CR, cs.CY

arXiv:2608.07471v1 Announce Type: new Abstract: This study considers the task of applying artificial intelligence to recognize bank fraud. In recent years, due to the COVID19 pandemic, bank fraud has become even more common due to the massive transition of many operations to online platforms and the...

📖 Read original article


2. Data-Driven Fire-Zone Segmentation for Improved Short-Term Wildfire Prediction ​

Author: Nicolas Caron, Christophe Guyeux, Hassan Noura, Benjamin Aynes
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07472v1 Announce Type: new Abstract: Wildfire prediction models typically discretize study areas into uniform grids, ignoring the heterogeneous spatial distribution of ignitions. We challenge this paradigm by showing that how data is discretized matters more than which model is used. We p...

📖 Read original article


3. Evolving Safety Landscape of Multi-modal Large Language Models: A Survey of Emerging Threats and Safeguards ​

Author: Xi Li, Shu Zhao, Xiaohan Zou, Fei Zhao, Fuxiao Liu, Yusen Zhang, Cheng Han, Yushun Dong, Jiaqi Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY

arXiv:2608.07535v1 Announce Type: new Abstract: Multi-modal large language models (MLLMs) integrate heterogeneous modalities through modality alignment and fusion, enabling stronger understanding and reasoning. However, this architectural shift reshapes the safety landscape of machine learning. Incr...

📖 Read original article


4. Tracing sources of epistemic uncertainty in deep learning predictions: homo- and hetero-scedastic linearized estimators ​

Author: Pierre Nodet, Thomas George
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.07630v1 Announce Type: new Abstract: We adapt two classical statistical estimators for quantifying uncertainty to modern deep learning, in order to provide clearer insights into uncertainty attributable to two sources : aleatoric uncertainty, or locally scarce data. Our approach leverages...

📖 Read original article


5. SkillConsist: Detecting Inconsistencies in Agent Skills via Bidirectional Graph Alignment ​

Author: Chaofan Meng, Yuhang Zheng, Yingnan Zhou, Sihan Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.07639v1 Announce Type: new Abstract: Agent Skills provide reusable capabilities to LLM agents. Agent Skill inconsistencies can expose undisclosed dangerous behavior or cause wrong Skill selection. Recent Agent Skill research has increasingly examined Agent Skill consistency detection. Exi...

📖 Read original article


6. PhysAttNet: Enhancing Predictive Performance in Industrial and Astrophysical Time Series via Physics-Informed Attention ​

Author: Amal Saadallah, Julia Tjus, Petra Wiederkeher, Wolfgang Rhode
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, physics.data-an

arXiv:2608.07681v1 Announce Type: new Abstract: Accurate and robust time series forecasting is essential in many applications involving physical processes, such as manufacturing monitoring and astrophysical event detection. In these settings, predictive models must remain reliable under noise, varia...

📖 Read original article


7. CODS: Iterative Bellman-Residual Data Selection for Reusable Offline Reinforcement Learning ​

Author: Ibne Farabi Shihab, Sanjeda Akter, Abu Sa-Adat Mohamed Moon-Im Al Ahsan, Md Najmus Swaqeeb, Anuj Sharma
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07719v1 Announce Type: new Abstract: Offline reinforcement learning repeatedly trains policies from a fixed transition pool, making redundant data costly across seeds and hyperparameters, while naive subsampling can remove rare transitions needed for long-horizon credit assignment. We int...

📖 Read original article


8. Neural Operators for Immersed-Boundary Soft Swimmers Locomotion ​

Author: Mohammad Sadegh Eshaghi, Yizheng Wang, Navid Valizadeh, Xiaoying Zhuang, Timon Rabczuk
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn

arXiv:2608.07722v1 Announce Type: new Abstract: High-fidelity immersed-boundary simulation resolves the coupled motion of a deforming swimmer and its surrounding flow, but the resulting cost limits repeated evaluations for engineering design, parameter studies, and control. We develop neural-operato...

📖 Read original article


9. Finite Constant Frontiers and Auditable Regret Certificates for Average-Reward Reinforcement Learning ​

Author: Ibne Farabi Shihab, Abu Sa-Adat Mohamed Moon-Im Al Ahsan, Md Najmus Swaqeeb
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07725v1 Announce Type: new Abstract: Average-reward reinforcement-learning regret is known up to logarithmic factors, but the numerical content of published guarantees is difficult to compare because probability mode, structural parameter, logarithmic normalization, prior information, and...

📖 Read original article


10. LUCID: Latent-Skill Unified Control via Imagined Dynamics for Long-Horizon Humanoid Loco-Manipulation ​

Author: Cheng Guo, Mingzhe Ni, Angelo Cangelosi, Arash Ajoudani
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.07746v1 Announce Type: new Abstract: Long-horizon humanoid loco-manipulation requires composing versatile whole-body skills and reliable high-level decision making. Existing methods often coordinate pretrained skills with scripted planners, finite-state machines or task-specific model-fre...

📖 Read original article


11. From Benchmark Performance to Tool Deployment: Human-in-the-Loop Anomaly Detection ​

Author: Mike Szklarzewski, CJ George, Gavin Smithson, Christopher Stokes, Dakota Fulp, William M. Jones, Benjamin Wynn, Alexander Ur, Agit Yesiloz, Clint Kallenbach, Mark Swartz, Nathan DeBardeleben, Sharmistha Chakrabarti
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.07770v1 Announce Type: new Abstract: Automated anomaly detection methods often report strong performance on curated academic benchmarks, but their behavior under real-world industrial conditions is less clear. In this work, we evaluate 19 unsupervised anomaly detection models on the BowTi...

📖 Read original article


12. The Sample Complexity of Policy Learning with Mu-Resets ​

Author: Gene Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07772v1 Announce Type: new Abstract: We study policy-based reinforcement learning under the $\mu$-resets interaction protocol of Kakade and Langford [KL02]. This interaction protocol enables the learner to sample trajectories from a given exploratory reset distribution $\mu$, in addition ...

📖 Read original article


13. Shape Mutating Expert Compression:LorExperts and BTExperts ​

Author: Inesh Chakrabarti, Sourjya Roy, Bowen Bao, Thiago Crepaldi, Spandan Tiwari, Ashish Sirasao
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.07814v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) language models deliver high capacity at low per-token compute, but deploying them cheaply requires compressing their many expert weight matrices. Expert pruning (e.g., REAP) and merging reduce cost but sacrifice accuracy and r...

📖 Read original article


14. From token probabilities to calibrated confidence: An empirical study of mathematical question answering ​

Author: Avery Ma, Lorne Schell, Vin Bhaskara, Leila Pishdad
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.07827v1 Announce Type: new Abstract: Confidence estimation for large language models (LLMs) aims to estimate the probability that a generated answer is correct, while calibration aligns these estimates with empirical accuracy. Prior work has shown that token probabilities are often overco...

📖 Read original article


15. TEMPER: Tensorized Efficient Manifold-constrained Parameterization for Expressive Residual Routing ​

Author: Yuxuan Gu, Wuyang Zhou, Huijun Xing, Danilo Mandic
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.07851v1 Announce Type: new Abstract: Residual connections rely on a static residual pathway, and are essential for training deep neural networks. Hyper-connections (HC) increase the expressivity of residual routing by incorporating multiple residual streams and learning dynamic informatio...

📖 Read original article


16. CommitKV: Lifecycle-Aware KV Cache Compression via Commit Transitions for Multi-Turn Agents ​

Author: Weizhong Huang, Jinchao Zhang, Xiawu Zheng
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07855v1 Announce Type: new Abstract: Multi-turn Reasoning-and-Acting (ReAct) agents accumulate growing trajectories of reasoning, tool calls, and observations. Their key-value (KV) caches grow accordingly, increasing memory use and attention cost during model inference. Existing KV cache ...

📖 Read original article


17. Adaptive KappaSharp: Condition-Number Shaping for Preferential Bayesian Optimization ​

Author: Ketong Shao, Jialu Wang, Xuekai Pei, Ali Mesbah
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.07859v1 Announce Type: new Abstract: Preferential Bayesian optimization (PBO) optimizes objectives accessible only through pairwise user comparisons. The standard approach fits a Gaussian process surrogate for observed pairwise comparisons (PairwiseGP) using the Laplace approximation and ...

📖 Read original article


18. CONFER: Conflict-Aware Evidence Negotiation for Regime-Calibrated Weak Supervision in Multimodal Emotion Recognition ​

Author: Bojing Hou, Ruohao Li, Yitong Zhu, Luwen Yu, Yuyang Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07867v1 Announce Type: new Abstract: Multimodal emotion recognition often treats self-reported labels as reliable supervision while overlooking self-report unreliability and cross-modal conflict. We propose \textbf{CONFER}, a graph-based conflict-aware evidence negotiation framework for w...

📖 Read original article


19. V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control ​

Author: Donghu Kim, Youngdo Lee, Hojoon Lee, Johan Obando-Ceron, Byungkun Lee, Aaron Courville, Pablo Samuel Castro, Jaegul Choo, Clare Lyle
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.07870v1 Announce Type: new Abstract: Improving sample efficiency remains a core challenge in reinforcement learning (RL), especially in real-world settings like robotics, where data collection is costly. This challenge is pronounced in visual RL, where high-dimensional inputs often obscur...

📖 Read original article


20. Router Sensitivity Under Lightweight Fine-Tuning Identifies Prunable Experts in Mixture-of-Experts Models ​

Author: Ali Janati, Kaoutar El Maghraoui, Xinyi Luo, Wenyuan Shen, Owen Zou, Yankai Mao
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.07890v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models decouple total parameters from per-token compute, but deployment still requires storing every expert. Recent theory shows that pruning experts with the smallest router-norm changes during fine-tuning can preserve accurac...

📖 Read original article


21. LLM-Based Embeddings for Program Analysis and Optimization ​

Author: Calvin Higgins, Marco Alvarez
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.PL

arXiv:2608.07894v1 Announce Type: new Abstract: Recent advances have highlighted the potential of machine learning, particularly Large Language Models (LLMs), for analyzing and optimizing programs. We present the first application of program embeddings from LLMCompiler---an LLM massively pretrained ...

📖 Read original article


22. When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes ​

Author: Yu Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.PF

arXiv:2608.07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache management an attractive lever: a policy that raised the hit rate would cut expert traffic per token...

📖 Read original article


23. SPECTRA: Pushing the KV Cache Beyond the 2-Bit Cliff via Spectral Transform Coding ​

Author: Jiamu Zhang, Liang Wu, Kelly Wan, Hanjie Chen, Liangjie Hong
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07915v1 Announce Type: new Abstract: Large language models (LLMs) increasingly read long inputs in the agentic era, from whole documents and codebases to conversations across many turns. Their inference memory is then dominated by the key-value (KV) cache, the stored attention keys and va...

📖 Read original article


24. Spectral Outliers Reveal Dominant Learned Structure in Transformer Attention ​

Author: Kasun Dewage, Marianna Pensky, Suranadi De Silva, T. H. Bandara
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.07921v1 Announce Type: new Abstract: We apply Marchenko-Pastur (MP) random matrix theory to pre-trained attention weights in order to separate each projection matrix into a random-like bulk and a set of spectral outliers. We validate this decomposition causally: zeroing the MP-identified ...

📖 Read original article


25. Information Routing across Batch Boundaries: Memory--Batch Tradeoffs in Lipschitz Bandits ​

Author: Zicheng Lyu, Zengfeng Huang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.07922v1 Announce Type: new Abstract: Adaptive learning needs both a state that preserves what observations imply and opportunities to act on that state. We study this width--depth tradeoff in stochastic Lipschitz bandits. After each pull, the learner retains at most $W$ bits of live rewar...

📖 Read original article


26. Second Order Drifting Models ​

Author: Drake Brown, Yuhao Huang, Shih-Hsin Wang, Bao Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NA, math.NA

arXiv:2608.07924v1 Announce Type: new Abstract: Drifting models are a recent class of one-step generative models that evolve the model distribution during training using a predefined sample-based drift field. Although they avoid iterative inference, their kernel-based drift fields induce frequency-d...

📖 Read original article


27. Adaptive Supervised Anchoring for On-Policy Self-Distillation ​

Author: Meilin Yang (Renmin University of China, Beijing, China), Zixuan Ding (Renmin University of China, Beijing, China), Jianhao Nie (Renmin University of China, Beijing, China), Weite Zhang (Renmin University of China, Beijing, China), Yuxin Zhang (Renmin University of China, Beijing, China), Zhiming Shao (Renmin University of China, Beijing, China), Li Yu (Renmin University of China, Beijing, China), Zhe Fu (Renmin University of China, Beijing, China)
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07935v2 Announce Type: new Abstract: On-policy self-distillation (OPSD) adapts a language model by distilling guidance from a frozen teacher on trajectories sampled from the student. Its effectiveness, however, depends critically on the quality of those trajectories. We show that when stu...

📖 Read original article


28. Persistent Semantic Entities in Tool-Augmented LLM Systems ​

Author: Zhaohui Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR, cs.SE

arXiv:2608.07952v1 Announce Type: new Abstract: Tool-augmented LLM agents can harbor implicit state that persists across sessions, activates through events, and propagates across agent boundaries---largely invisible to standard debugging. We formalize this as Persistent Semantic Entities (PSEs): con...

📖 Read original article


29. From Uncertainty to Failure Attribution: Self-Diagnosing Models for Failure Attribution under Distribution Shift ​

Author: Yiyao Yang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07953v1 Announce Type: new Abstract: Distribution shift poses a significant challenge to the robustness of machine learning models, but the current solutions only aim to detect out-of-distribution (OOD) samples and predict uncertainty levels. We introduce a problem setting for failure att...

📖 Read original article


30. EasyBalance: Cross-Layer Load Balancing in Distributed MoE Inference ​

Author: Yize Wu, Ke Gao, Ling Li, Yanjun Wu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.07964v1 Announce Type: new Abstract: Load Balancing has emerged as a critical problem in expert-parallel distributed inference of Mixture-of-Experts (MoE) models. As routing distributions are typically skewed across experts, devices hosting lighter-loaded experts must idle to wait for the...

📖 Read original article


31. ZeroLock: Concurrent Memory-Efficient LLM Training via Modular Update Decoupling ​

Author: Wentao Dai, Xuanran Li, Yuxiang Zhang, Ming Tang, Chao Huang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2608.07974v1 Announce Type: new Abstract: Large language model (LLM) fine-tuning at the edge adapts the model to scenario-specific data while preserving privacy. Although existing studies proposed pipeline parallelism to address the limited memory and computing resources of edge devices, they ...

📖 Read original article


32. Evaluator Ensembles Under Reward Hacking: Covariance Geometry and Finite-Search Guarantees ​

Author: Fariya Afrin, Ibne Farabi Shihab
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08002v1 Announce Type: new Abstract: Language-model judges and reward models enable scalable supervision, but finite optimization can exploit evaluator errors rather than improve response quality. We characterize this failure through the covariance geometry of evaluator ensembles. For cal...

📖 Read original article


33. Quality-Diversity Stress Tests for Process Reward Models:What Archive Coverage Can and Cannot Certify ​

Author: Ibne Farabi Shihab, Fariya Afrin
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08008v1 Announce Type: new Abstract: Process reward models (PRMs) score intermediate reasoning steps and are widely used for search, ranking, and training, but optimization can exploit these learned proxies by increasing reward while turning correct reasoning into incorrect reasoning. We ...

📖 Read original article


34. Ground-Truth Neighborhood Regularization for Reinforcement Learning Post-Training of Time Series Foundation Models ​

Author: Jianqi Zhang, Xingyu Zhang, Zeen Song, Changwen Zheng, Fanjiang Xu, Wenwen Qiang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08010v1 Announce Type: new Abstract: Time series forecasting (TSF) plays an important role in a wide range of real-world applications. Recently, time series foundation models (TSFMs), pretrained on large-scale datasets, have demonstrated strong generalization capabilities and emerged as a...

📖 Read original article


35. Do All LLMs Know When They're Being Harmful? A Reproducibility Study of Latent-Space Safety Probes Across Model Families ​

Author: Alizishaan Khatri, Dun Li Chan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2608.08029v1 Announce Type: new Abstract: Khatri et al. (2026) [DOI: 10.1109/DSN-W70714.2026.00027] show that lightweight MLP probes on final-layer activations of a single 8B model (LLaMA-3.1-8B) detect harmful prompts at F1 competitive with guard models 1000x larger, using one probe per bench...

📖 Read original article


36. CLAM: Causal Spatial Disaggregation to Infer Local Effects From Coarse Data ​

Author: Gerrit Gro{\ss}mann, Sumantrak Mukherjee, Sebastian J. Vollmer
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08064v1 Announce Type: new Abstract: Learning fine-grained spatial patterns from coarse-resolution data is challenging, especially in causal settings where high-resolution effects must be inferred from aggregated interventions and outcomes. We introduce CLAM, a method for estimating local...

📖 Read original article


37. Adaptive Symmetry Discovery for Dynamical System Identification ​

Author: Behrooz Tahmasebi, Melanie Weber
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DS

arXiv:2608.08091v1 Announce Type: new Abstract: Dynamical systems model trajectory data generated by fixed underlying dynamics, with applications ranging from biology to physics. Especially in scientific settings, dynamical systems are not generic but often exhibit symmetries imposed by physical law...

📖 Read original article


38. Support Selection Beyond Smooth DAG Exactness: Completion Geometry,Score Margins, and Selective Certificates ​

Author: Rui Wu, Zongyuan Chen, Hong Xie
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08103v1 Announce Type: new Abstract: Smooth acyclicity constraints answer whether a weighted support is a DAG, whereas structure learning asks which support change should be made. Existing analyses establish degeneracy for particular constraint formulas but do not isolate what follows fro...

📖 Read original article


39. TSDS-Toolbox: A Toolbox for Measuring Time-Series Dataset Similarity ​

Author: Yen-Ku Liu, Hongjie Chen, Ryan A. Rossi, Franck Dernoncourt
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08119v1 Announce Type: new Abstract: The rapid advancement of artificial intelligence (AI) has significantly accelerated research in time-series analysis, particularly in forecasting, classification, and generation tasks. Recent models, especially foundation models, benefit from time-seri...

📖 Read original article


40. Accurate Ensembles, Fragile Narratives: Multi-Scale Stacking and a Fidelity Audit of LLM-Generated Explanations for Credit Risk ​

Author: Gregorius Reynaldi Pratama, Kuo-Kun Tseng
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.08126v1 Announce Type: new Abstract: Credit scoring increasingly relies on models whose decision logic cannot be read off their parameters, in tension with supervisory expectations that adverse decisions be explainable. A common proposal closes that gap with a language model: compute feat...

📖 Read original article


41. DoGMA: A Central-Dogma-Guided Foundation Model for Multi-Omics Alignment and Multi-Task Learning in Oncology ​

Author: Junfei Ling (Institute of Medical Robotics, Shanghai Jiao Tong University), Bangzheng Pu (Institute of Medical Robotics, Shanghai Jiao Tong University), Bingsen Xue (Institute of Medical Robotics, Shanghai Jiao Tong University), Tianle Li (Institute of Data Science, The University of Hong Kong), Ruying Hu (Oriental Pan-Vascular Devices Innovation College, University of Shanghai for Science and Technology), Cheng Jin (Institute of Medical Robotics, Shanghai Jiao Tong University)
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08148v1 Announce Type: new Abstract: Attention mechanisms have been widely utilized in modern deep learning, and many existing multi-omics models inherit their conventional use to allow unrestricted bidirectional interactions. However, the fundamental logic of life is directional. Existin...

📖 Read original article


42. A Hybrid Nested Harness for Decoupling Structure and Parameters in LLM-Driven Optimization ​

Author: V'ictor Gallego
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2608.08156v1 Announce Type: new Abstract: In evolutionary algorithms powered by language models, the LLM acts as a single operator that simultaneously updates structural components (like control flow) and continuous parameters. While LLMs can be good at the first, they are not efficient at the...

📖 Read original article


43. Predicting blood clot growth from sparse post-onset measurements with latent neural differential equations ​

Author: Lennon J. Shikhman, Ying Qian, He Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM, q-bio.TO

arXiv:2608.08165v1 Announce Type: new Abstract: Computational models of blood clotting improve understanding of thrombus formation, but their clinical application remains limited because many model inputs are difficult to measure and patient-specific data are often sparse. We present a computational...

📖 Read original article


44. Biologically Informed Representation Learning for Robust Cross-Center Generalization of MALDI-TOF Mass Spectrometry ​

Author: Alejandro L. Garc'ia-Navarro, Carlos Sevilla-Salcedo, Bel'en Rodr'iguez-S'anchez, Vanessa G'omez-Verdejo
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08182v1 Announce Type: new Abstract: Machine learning models for MALDI-TOF mass spectrometry have shown considerable promise for clinical microbiology tasks such as microbial identification and antimicrobial resistance prediction. However, their deployment across institutions remains limi...

📖 Read original article


45. Beyond Aggregate Calibration: Decomposing Income-Conditional Recall Disparities in Automated Credit Default Prediction ​

Author: Sai Srikar Boddupalli
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2608.08202v1 Announce Type: new Abstract: Data-centric curation pipelines frequently rely on model confidence scores to flag and filter noisy or mislabeled training instances. Evaluating this filtering convention on a large-scale consumer lending sample (LendingClub, N = 1,344,936) uncovers an...

📖 Read original article


46. FreSH: Frequency-Segmented Hierarchical Multi-Expert Framework for Multivariate Time Series Classification ​

Author: Pingping Liu, Muyao Wang, Zijian Zhang, Tongshun Zhang, Hao Miao, Guorui Xie, Qingliang Li, Qiuzhan Zhou
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08207v1 Announce Type: new Abstract: Multivariate Time Series Classification (MTSC) demands models that can effectively capture complex temporal patterns across multiple scales while remaining computationally efficient. However, existing approaches generally struggle to reconcile fine-gra...

📖 Read original article


47. Control-Diverse Reinforcement Fine-Tuning: Decoupling the Shared Control Bottleneck of RL Post-Training ​

Author: Binwen Tan, Jingchao Wang, Dengzhe Hou, Lingyu Jiang, Zeyuan Wu, Yunhan Shen, Fangzhou Lin, Kazunori Yamada, Atsushi Koike
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08224v1 Announce Type: new Abstract: Reinforcement learning post-training unlocks complex reasoning in LLMs. Yet benchmark scores reveal only whether a model improved, not what changed inside it, nor how it splits finite capability across tasks. A representative interpretability line attr...

📖 Read original article


48. SAGE: SLO-Aware Adaptive Retrieval for Production RAG Systems ​

Author: Muhammad Faizan Raza (Luna), Shuo (Luna), Yang, Satish Mahadevan Srinivasan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.DC, cs.IR

arXiv:2608.08237v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems in production operate under strict service level objectives (SLOs) on tail latency and infrastructure cost. However, standard retrieval pipelines rely on fixed retrieval budgets that ignore query difficulty,...

📖 Read original article


49. The Replay Gap: Static Evaluation of Model Switching in LLM Agents Scores the Wrong World ​

Author: Ashritha Gonuguntla
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.08239v1 Announce Type: new Abstract: LLM routers promise efficiency by matching each request to the cheapest adequate model, and are increasingly applied per step inside multi-step agents. Yet agentic routers are evaluated like single-turn routers: by replaying logged trajectories and sub...

📖 Read original article


50. Learning from Environmental Feedback: Credit Assignment across Multiple Timescales for Agentic Reinforcement Learning ​

Author: Yifu Huo, Shunjie Xing, Chenglong Wang, Peinan Feng, Qiaozhi He, Yan Ding, Anxiang Ma, Yuxin Gao, Tongran Liu, Tong Xiao, Jingbo Zhu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.08255v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) often suffers from delayed and sparse rewards in real-world environments. A promising solution to this challenge is credit assignment, which aims to decompose trajectory-level rewards and provide more fine-grained su...

📖 Read original article


51. Opportunity Is Not Realizability: Selection-Valid Diagnostics for Multi-LLM Routing ​

Author: Ibne Farabi Shihab, Abu Sa-Adat Mohamed Moon-Im Al Ahsan, Md Najmus Swaqeeb
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08265v1 Announce Type: new Abstract: Oracle routing measures how much a pool of language models could gain from per-query selection, but the diagnostic has two flaws: testing against a best fixed model selected on the same examples invalidates paired inference, and a full-information orac...

📖 Read original article


52. Stateful CARS: Exact Cross-History Reuse for Policy-Constrained LLM Agents ​

Author: Ibne Farabi Shihab, Md Najmus Swaqeeb, Abu Sa-Adat Mohamed Moon-Im Al Ahsan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08282v1 Announce Type: new Abstract: Tool-using language-model agents face constraints whose meaning changes with observations and prior actions. We study exact sampling from the model distribution conditioned on a hard stateful validator while reusing invalidity certificates across histo...

📖 Read original article


53. Causal State-Space Model for Causal Inference: Estimating Longitudinal Individual Treatment Effects ​

Author: Abisoye Abidakun, Mingjun Zhong, Georgios Leontidis
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.08288v1 Announce Type: new Abstract: Estimating counterfactual outcomes over time from longitudinal observational data is central to clinical decision support. Existing methods rely on domain confusion -- adversarial training that renders representations invariant to treatment assignment ...

📖 Read original article


54. A Controlled Study of Feature-Based Knowledge Distillation Across Student Designs ​

Author: Abhinand Balachandran, Praveen Prashant
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.08294v1 Announce Type: new Abstract: Knowledge distillation trains a smaller student to match the outputs of a larger teacher. Feature-based methods also align intermediate representations, but this extra constraint may affect students differently. We study this question on CIFAR-100 usin...

📖 Read original article


55. Machine-Learning-Based Diagnostic Framework for Passive Ultrasonic Detection of Railway Wheel Defects ​

Author: Aashish Shaju, Steve Southward, Mehdi Ahmadian
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.RO, eess.SP

arXiv:2608.08301v1 Announce Type: new Abstract: Reliable identification of railway wheel defects is important for safety and maintenance. This study develops a machine-learning-based diagnostic framework for multi-class defect identification using passive air-coupled ultrasonic acoustic emission sig...

📖 Read original article


56. The Neural Division of Labor: Biologically-Inspired Modular Architectures for Robust Neuromorphic Computing ​

Author: Maksim Bazhenov, Serafim Grubas, Vakhtang Putkaradze
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08317v1 Announce Type: new Abstract: Biological neural systems achieve high efficiency and robustness through compartmentalized architectures. In contrast, modern artificial neural networks rely on globally entangled structures, which obscure decision logic and suffer from catastrophic fo...

📖 Read original article


57. Spatial Heterogeneity-Aware Multi-Hazard Susceptibility and Risk Mapping at Regional Scale ​

Author: Aswathi Mundayatt, Siddharth Anil, Hitanshu Seth, Jaya Sreevalsan-Nair
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08321v1 Announce Type: new Abstract: Floods and landslides often co-occur, but their relationships with environmental controls vary spatially. This study develops a spatial heterogeneity-aware framework for flood-landslide susceptibility and relative-risk mapping in Kerala, India, and Nep...

📖 Read original article


58. PRISM: A Predictive Protocol for Permutation Optimization via Landscape Diagnostics ​

Author: Blessings Mambwe
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC

arXiv:2608.08344v1 Announce Type: new Abstract: Permutation optimization arises whenever the components of a system are fixed but their ordering affects performance. We introduce PRISM, a predictive protocol for permutation optimization that measures a fitness landscape before selecting a search str...

📖 Read original article


59. Correlation flow governs learning at criticality ​

Author: Andrea Combette, Nelly Pustelnik, Antoine Venaille
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn

arXiv:2608.08350v1 Announce Type: new Abstract: The initialisation of deep neural networks determines whether information and gradients can propagate across depth, yet a unified theory connecting these properties to learning dynamics remains elusive. Combining mean-field theory and random matrix the...

📖 Read original article


60. Unimodality-Promoting Regularized Learning for Ordinal Regression ​

Author: Ryoya Yamasaki
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08359v1 Announce Type: new Abstract: Ordinal regression, also called ordinal classification, is classification of ordinal data, in which the underlying target variable is categorical and considered to have a natural ordinal relation. Previous works have indicated that, in many real-world ...

📖 Read original article


61. Exact Rank and Convex Calibration Dimension Lower Bounds for the Multi-Label F1 Loss ​

Author: Mingyuan Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.08399v1 Announce Type: new Abstract: The instance-wise $F_1$ measure is a central performance measure for multi-label classification. For a problem with $s$ labels, it defines a $2^s\times 2^s$ loss matrix. Previous work exhibited $s^2+1$-coordinate affine and shifted low-rank representat...

📖 Read original article


Author: Yiqiao Liao, Parinaz Naghizadeh
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08406v1 Announce Type: new Abstract: Existing learning-based influence maximization frameworks rely heavily on complex neural architectures and continuous optimization over seed representations. We challenge this paradigm with SIMBA, a diffusion-model-agnostic framework pairing a lightwei...

📖 Read original article


63. Constrained Learning with Universally Learnable Concept Classes ​

Author: Herlock SeyedAbolfazl Rahimi, Spyridon Pougkakiotis, Dionysis Kalogerias
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, math.OC, math.ST, stat.ML, stat.TH

arXiv:2608.08414v1 Announce Type: new Abstract: We study constrained statistical learning over infinite-dimensional hypothesis classes in the fully nonconvex setting, and establish universal PACC learnability of the solutions of dual algorithms: Probably Approximately Correct on Constraints, guarant...

📖 Read original article


64. Optimal Learning Under Tsybakov Noise ​

Author: Steve Hanneke, Hongao Wang, Mingyue Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08416v1 Announce Type: new Abstract: Probably Approximately Correct (PAC) learning [Val84] is a fundamental learning model that has been extensively investigated. In this model, $\mathcal{H} \subseteq {0,1}^{\mathcal{X}}$ is a concept class, and $h^*\in\mathcal{H}$ is the target concept...

📖 Read original article


65. FSTC-Encoder: Feature--Spatial--Temporal Correlation Learning for Generalizable RF Sensing ​

Author: Jing Wang, Zhu Wang, Changlong Cheng, Yifan Guo, Yin Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08439v1 Announce Type: new Abstract: Heterogeneous RF sensing differs substantially in feature structure, spatial layout, and temporal scale, making existing models difficult to reuse across devices, environments, and RF modalities. We propose FSTC-Encoder, which unifies heterogeneous RF ...

📖 Read original article


66. MGMCL: Multi-Granularity Manifold Contrastive Learning With Neural ODEs for Cross-Subject EEG Emotion Recognition ​

Author: Xiang Xie
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08440v1 Announce Type: new Abstract: Cross-subject electroencephalogram (EEG)-based emotion recognition remains challenging due to substantial inter-individual variability and discrete formulation that overlooks affective continuity. Existing methods operate in Euclidean space and focus o...

📖 Read original article


67. No Unique Minimizer, No Problem: On the Consistency of Robust Neural Classifiers ​

Author: Subhabrata Majumdar, Anand Deo, Partha Pratim Saha, Abhik Ghosh
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2608.08489v1 Announce Type: new Abstract: Neural network classifiers trained by cross-entropy minimization are highly sensitive to label noise and adversarial contamination. While robust alternatives offer bounded influence and resistance to corruption, their statistical foundations in the dee...

📖 Read original article


68. Out-of-Distribution Federated Distillation with Domain-Aware Proxy ​

Author: Jiahao Xiao, Jiangming Liu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08525v1 Announce Type: new Abstract: Federated Learning is a distributed machine learning paradigm that trains a global model by aggregating local clients without sharing private data of each client. Federated Distillation (FD) builds upon this paradigm by leveraging knowledge distillatio...

📖 Read original article


69. Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing ​

Author: Srinivasan Manoharan, Junhua Zhao, Fangbo Tu, Haifeng Wu, Jian Wan, Maliah Rajan M, Ashwin Hegde, Mithun Sasidharan, Kalyan Chakravarthi Podamekala
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08528v1 Announce Type: new Abstract: Enterprise AI coding assistants incur substantial inference spend, and naive token-cost minimization often fails to reduce end-to-end cost once retries, escalations, and developer wait time are included. We present Task-to-Model Optimization (T2MO), a ...

📖 Read original article


70. Can Graph Learning Learn Circuits? ​

Author: Chester Tan, Moritz Lampert, Courtney Maynard, Ankit Ramakrishnan, Tina Eliassi-Rad, Ingo Scholtes
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08536v1 Announce Type: new Abstract: Circuit localization is a mechanistic interpretability task whose goal is to identify a sparse subgraph of a transformer's computation graph sufficient to reproduce a particular behavior. Most established methods localize circuits independently for eac...

📖 Read original article


71. When Skills Meet Safety: Benchmarking and Characterizing the Adaptive Jailbreak Robustness of Skill-Merged LLMs ​

Author: Yu Ma, Hongli Shi, Jing Li, Xinran Xu, Weiwei Hou
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08542v1 Announce Type: new Abstract: Model merging has become the default way to give an aligned language model new skills without retraining: a practitioner folds task vectors from math, code, or domain specialists into a safety-aligned base using task arithmetic, TIES, or DARE. This con...

📖 Read original article


72. Neural Message Passing on Structural Interaction Graphs for Fully-Inductive Graph Neural Networks ​

Author: Omer Yom Tov, Avigdor Gal
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08567v1 Announce Type: new Abstract: A central obstacle in building graph foundation models is the input heterogeneity in terms of feature space dimensionality, semantics, and structure. Such heterogeneity limits the capability of graph neural networks to generalize to new graphs with uns...

📖 Read original article


73. Robust Reputation-Driven Crowdsourced Federated Learning ​

Author: Mouhamed Amine Bouchiha, Gregory Blanc
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.DC

arXiv:2608.08574v2 Announce Type: new Abstract: Crowdsourced Federated Learning (CrowdFL) extends traditional federated learning by enabling open and heterogeneous participation through a crowdsourcing paradigm. In this setting, reputation-driven incentive mechanisms are commonly employed to guide w...

📖 Read original article


74. When Can Fraud Operations Authorize Automation? A Decision-Support Framework for Fresh Audit Evidence and Review Workload ​

Author: Jie Deng (Tongji University, Shanghai, China)
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.08577v1 Announce Type: new Abstract: Fraud operations must allocate events among automatic approval, analyst review, and automatic blocking even though the labels needed to evaluate these actions are selective and delayed. Predictive scores order cases, but they do not show whether the ev...

📖 Read original article


75. Multi-Agent Reinforcement Learning via Agent-Specific Preference ​

Author: Ni Mu, Yao Luan, Yiqin Yang, Qing-Shan Jia
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08604v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is a powerful framework for solving complex collaborative tasks, but it relies heavily on well-defined global reward functions. Designing such rewards is challenging, especially in systems with heterogeneous ag...

📖 Read original article


76. Domain-Aware Pruning: Sparsity and Domain Generalization via Regularized Probabilistic Masking ​

Author: Parham Sazdar, Mostafa Tavassolipour, Reshad Hosseini
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08624v1 Announce Type: new Abstract: Domain generalization (DG) and neural network pruning are conventionally treated as distinct objectives, targeting out-of-distribution (OOD) robustness and model efficiency, respectively. In this work, we bridge this gap by introducing Domain-Aware Pru...

📖 Read original article


77. Trajectory Design and Budgeted Querying for Digital Twin Calibration ​

Author: Vladyslava Spitkovska, Dmytro Kuzmenko
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2608.08631v1 Announce Type: new Abstract: Digital-twin calibration requires interaction data that is expensive to collect. We study two acquisition decisions: which trajectories to generate, and when to spend a limited budget on privileged parameter measurements. Our framework couples an excit...

📖 Read original article


78. Exact Rank-Space KL Projection for Shared-Marginal Low-Rank Factors: Application to Doubly Stochastic Clustering ​

Author: Enliang Hu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.08642v1 Announce Type: new Abstract: We study exact Kullback--Leibler (KL) projection for low-rank factorizations whose two nonnegative factors have prescribed row marginals and a shared, learned column marginal. For arbitrary positive row marginals of equal total mass, the joint KL proje...

📖 Read original article


79. Path-dependent Discrete Amortized Inference ​

Author: Tiago da Silva, Esmeralda S. Whitammer, Salem Lahlou
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08644v1 Announce Type: new Abstract: We consider the problem of sampling compositional and discrete objects from a given unnormalized posterior distribution. Notably, recent studies have shown that this problem can be efficiently solved by learning a deterministic Markov Decision Process ...

📖 Read original article


80. Multi-Relational Knowledge Graph Enhanced Embedding for Trajectory-User Linking ​

Author: Zhifeng Chu, Bin Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.08646v1 Announce Type: new Abstract: Trajectory-User Linking (TUL) aims to identify the owner of an anonymous trajectory from a set of candidate users, providing a basis for user mobility analysis and personalized location-aware services. Existing methods often learn Point of Interest (PO...

📖 Read original article


81. LegoLM: Structured Weight Sharing for Large Language Models ​

Author: Joseph Bingham
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08652v1 Announce Type: new Abstract: We present \LegoLM{}, a structured weight-sharing compression framework for large language models grounded in a systematic study of why global weight sharing fails and how to fix it. We identify two distinct failure modes. Distributional mismatch: for ...

📖 Read original article


82. Catastrophic Forgetting in Continual Reinforcement Learning ​

Author: Emma Graham
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08673v1 Announce Type: new Abstract: This work explores the relationship between task similarity and catastrophic forgetting in reinforcement learning. Catastrophic forgetting, the phenomenon in machine learning of losing the ability to effectively perform on previous tasks, is a signific...

📖 Read original article


83. Backward Compatibility in Tree-Based Explanations and Enhanced CART Algorithm ​

Author: Hirofumi Suzuki
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08674v1 Announce Type: new Abstract: In the operation of machine learning models, model update is a fundamental process that requires careful consideration of its impact on downstream decision-making. Particularly when operating explainable models, changes in explanations resulting from m...

📖 Read original article


84. Efficient Test-Time Scaling for LLM-based Time Series Forecasting ​

Author: Xuan-May Le, Minh-Tuan Tran, Ling Luo, Uwe Aickelin, Dinh Phung, Trung Le
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08675v1 Announce Type: new Abstract: Long-term time series forecasting benefits from preserving global structure such as trends and seasonality. Recent LLM-based forecasters often improve accuracy through test-time scaling (e.g., iterative refinement), but these methods are computationall...

📖 Read original article


85. RippleKV: Cross-Layer KV Cache Allocation via Perturbation Propagation ​

Author: Dongjie Xu, Kai Qian, Julius, Weijie Shi, Yuxuan Sun, Minghua Tang, Fenglei Jin, Hanchi Dong, Jiajie Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08684v1 Announce Type: new Abstract: Long-context LLM inference is bottlenecked by KV cache memory, yet distributing a limited cache budget across layers remains challenging. Existing methods rely on proxies such as layer depth, attention statistics, or representation change. These proxie...

📖 Read original article


86. Loss-Resilient Wireless Video Token Communication over Block Fading Channels ​

Author: Bingyan Xie, Yongjeong Oh, Zihan Chen, Jihong Park, Yongpeng Wu, Wenjun Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.MM

arXiv:2608.08698v1 Announce Type: new Abstract: Video token communication represents video content as discrete tokens that differ in their importance to reconstruction and exhibit temporal dependencies. When these tokens are packetized for wireless transmission, block fading can cause multiple impor...

📖 Read original article


87. Gaming Without an Attacker: Benchmark Fingerprinting in LLM-Driven Search Under Selection Pressure ​

Author: V'ictor Gallego
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08722v1 Announce Type: new Abstract: Benchmarks for systems that are optimized against the evaluation signal measure something different from what they claim. We document this concretely in two GPU-kernel-optimization suites with held-out generalization gates: Metal-Sci (10 scientific-com...

📖 Read original article


88. PAST: Privileged Adaptation from Complete Student Trajectories for On-Policy Self-Distillation ​

Author: Yangyang Feng, Zhuoyan Feng, Junlan Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08726v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) uses a privileged teacher to supervise a reasoning model on prefixes sampled from its own rollouts. Yet each rollout also reveals how the student's response unfolds and whether it succeeds, student-specific hindsight ...

📖 Read original article


89. Measuring and Reducing WebGPU Dispatch Overhead for LLM Inference ​

Author: J\k{e}drzej Maczan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.PF

arXiv:2608.08730v2 Announce Type: new Abstract: Large Language Models are deployed to multiple types of environments, from internet browsers to edge devices, and WebGPU serves as a modern cross-platform standard. The engines for browser-based LLM inference have proliferated, yet the overhead of WebG...

📖 Read original article


90. Memory-Efficient Activation Checkpointing with Sliding Window and Hirschberg's Algorithm for 0/1 Knapsack Solving in PyTorch ​

Author: J\k{e}drzej Maczan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08740v1 Announce Type: new Abstract: Activation checkpointing minimizes the runtime of neural networks under a given memory budget, by selecting which intermediate tensors to store and which to recompute. PyTorch solves this as a 0/1 knapsack problem, where operations from a joint forward...

📖 Read original article


91. Learning from Consensus and Disagreement: Unsupervised On-Policy Self-Distillation with Minority-Trajectory Contrast ​

Author: Jiaxin Guo, Yanwei Yue, Xuanbo Fan, Chunyu Yang, Yan Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08764v1 Announce Type: new Abstract: On-policy self-distillation improves language-model reasoning by querying a teacher on states actually visited by the student. Recent methods create a powerful information asymmetry by exposing the teacher to privileged context, yet they fundamentally ...

📖 Read original article


92. Quantum-Classical Physics-Informed Kolmogorov-Arnold Networks for Solving Fuzzy Differential Equations ​

Author: Xiang Rao, Yuxuan Shen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08782v1 Announce Type: new Abstract: In this study, we propose a quantum-classical physics-informed Kolmogorov-Arnold network (QCPIKAN) dedicated to the solution of fuzzy differential equations. The network takes the spatiotemporal coordinates and membership level as joint inputs and empl...

📖 Read original article


93. Distilling Vision-Language Models for Robust Traffic Sign Perception in Autonomous Vehicles ​

Author: Pedram MohajerAnsari, Amir Salarpour, Mert D. Pes'e
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08815v1 Announce Type: new Abstract: Traffic sign recognition (TSR) models based on deep neural networks achieve strong clean-data performance but remain vulnerable to physically realizable adversarial attacks, including shadow perturbations, natural-light interference, and printed patche...

📖 Read original article


94. Hybrid Neural-Classical Correction for Frozen Time Series Foundation Models: A Comprehensive Ablation Study on High-Frequency Stock Prediction ​

Author: Kasun Dewage, Suranadi De Silva, Shankhadeep Mondal
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-fin.ST

arXiv:2608.08825v1 Announce Type: new Abstract: Foundation models for time series forecasting demonstrate impressive zero-shot generalization but often underperform on specialized domains such as high-frequency finance. We present a comprehensive study of hybrid neural-classical correction for adapt...

📖 Read original article


95. The Cost of Adaptivity: Matching Lower Bounds Across Learning Problems ​

Author: Ibne Farabi Shihab, Adria Binte Habib
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08826v1 Announce Type: new Abstract: Adaptive procedures must work without nuisance information an oracle may use, such as a gradient scale or smoothness index, and robust procedures may have to answer queries whose coordinate and inspection time are chosen only after the data are seen. S...

📖 Read original article


96. Beyond Routing: Decoupling Expert Dispatch and Aggregation in Sparse Mixture-of-Experts ​

Author: Zongfei Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08853v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) routers commonly use the same scores both to select experts and to weight their already-computed outputs. We study whether these two roles, dispatch and aggregation, should be coupled. On pretrained OLMoE-1B-7B, we keep ...

📖 Read original article


97. Agentic Anomaly Detection with ORCA-Style Dynamic Inductive Bias Adaptation in Multimodal Wearable Time Series Data ​

Author: Anushka Roy, Jyotirmoy Singh, Shreea Bose, Chittaranjan Hota
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.08859v1 Announce Type: new Abstract: Wireless Body Area Networks (WBANs) generate multivariate physiological time series that are highly nonstationary and must often be processed under strict computational and memory constraints. A critical yet underexplored challenge in this setting is s...

📖 Read original article


98. Approximation Rates for Metaplectic Neural Networks ​

Author: Ahmed Abdeljawad, Marcello Carioni, Elena Cordero
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, math.FA

arXiv:2608.08872v1 Announce Type: new Abstract: In this paper we develop quantitative approximation results for shallow neural networks constructed using a dictionary based on metaplectic operators. First, we extend the concept of Barron spaces by considering a symplectically motivated extension of ...

📖 Read original article


99. DistillCache: KL-Guided Adaptive KV-Cache Eviction for Memory-Efficient LLM Inference ​

Author: Asaad Althoubi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.PF

arXiv:2608.08878v1 Announce Type: new Abstract: Transformer-based large language models (LLMs) achieve strong performance across many tasks, but their Key-Value (KV) cache grows linearly with sequence length, creating a severe memory bottleneck for long-context inference. Existing heuristic eviction...

📖 Read original article


100. Federated Attention Autoencoders with a Stochastic Aggregation Scheme for Anomaly Detection ​

Author: Mihailo Ili'c, Milo\v{s} Savi'c, Vladimir Kurbalija, Mirjana Ivanovi'c, Giancarlo Fortino, Du\v{s}an Jakoveti'c
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08906v1 Announce Type: new Abstract: Outlier detection in decentralized data environments is a challenging task for many machine learning implementations, particularly in settings where data cannot be shared. Recently, there have been advances in federated outlier detection, some of which...

📖 Read original article


101. A Domain-Structured Ensemble Framework for Perioperative Outcome Prediction Using Electronic Health Record Data ​

Author: Shikhar Shukla, Cristina Barboi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08920v1 Announce Type: new Abstract: Perioperative risk prediction models are often limited by narrow surgical populations, incomplete intraoperative data, poor calibration, and limited interpretability. We present a domain-structured ensemble framework for perioperative outcome predictio...

📖 Read original article


102. Idea Search: Guiding Tree Search with Ideas to Explore Diverse Scientific Methods ​

Author: Xuefei Julie Wang, Hao Cui, Michael P. Brenner, Subhashini Venugopalan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.GN, q-bio.QM

arXiv:2608.08958v1 Announce Type: new Abstract: Tree Search-based test-time scaling of LLMs is a powerful tool for automated scientific coding. However, pure Tree Search sometimes struggles with systematic exploration, becoming trapped in local optima, or unproductive loops, especially in the vast s...

📖 Read original article


103. Gradient Under Microscope: Benchmarking Resource Utilization of Memory-Efficient Gradient Computation Methods ​

Author: Sarthak Mahapatra, Zihan Zhou, Khatoon Khedri, Mehdi Hosseinzadeh, Reza Rawassizadeh
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08961v1 Announce Type: new Abstract: AI training's rising resource intensity is straining electricity supplies and carbon budgets, motivating systematic study of memory-efficient training on constrained hardware. We benchmark five gradient optimizers (SGD, Adam, Adagrad, Adadelta, and Con...

📖 Read original article


104. Math-Vision Diagrams: A Comprehensive Benchmark for Evaluating LLM Mathematical Diagram Generation Capabilities ​

Author: Harish Kashyap, Kiran Byadarhaly, Sriram Chakaravarthy, Sanyukta Tuti, Aryan Mistry
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08964v1 Announce Type: new Abstract: The generation of mathematically precise diagrams from tex- tual prompts has emerged as a critical yet underexplored capability of Large Language Models (LLMs). This has been of interest to researchers in the areas of curriculum preparation, automated ...

📖 Read original article


105. Label-Free Parkinson's Disease Screening from Face and Voice through Mechanistic Interpretability ​

Author: Jiaheng Su, Yu Sun
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08976v1 Announce Type: new Abstract: Parkinson's disease (PD) is the second most common neurodegenerative disorder. Typical machine learning screening methods require PD labels, but the available data is limited by privacy concerns and the need for expert annotation. We propose a label-fr...

📖 Read original article


106. Twin Rollouts: Noise-Coupled Counterfactual Branching in Interactive Video World Models ​

Author: Yu Ma, Hongli Shi, Xinran Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08982v1 Announce Type: new Abstract: Interactive video world models generate rollouts autoregressively under an action stream, yet they are trained and evaluated almost exclusively on factual prediction. We study counterfactual generation inside the rollout: given a trajectory the model h...

📖 Read original article


107. SoftMCC: An MCC-Brier Calibration Bridge for Threshold-Free Model Selection under Class Imbalance ​

Author: "Ozkan Canay
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08984v1 Announce Type: new Abstract: Model selection for imbalanced binary classification often uses the Matthews correlation coefficient (MCC), but thresholding makes validation rankings threshold-dependent. SoftMCC is a post-training MCC validation framework on established probability-v...

📖 Read original article


108. Dynamic Distribution-Aware Uncertainty Tracking in Vision-Language Representation Learning ​

Author: Ao Zhou, Zhiwei Jiang, Zifeng Cheng, Cong Wang, Shufan Yang, Haoru Chen, Qing Gu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09011v1 Announce Type: new Abstract: Uncertainty Quantification (UQ) aims to measure the reliability of model predictions, serving as a critical safeguard for deploying Vision-Language Models (VLMs) in safety-critical scenarios. Post-hoc approaches are widely adopted due to their lightwei...

📖 Read original article


109. HOPPER: Learnable Hop Extraction for Linearized Graph Sequence Models ​

Author: Isuru Herath, Arin Gopakumar, Sharan Sahu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09031v1 Announce Type: new Abstract: Graph neural networks typically propagate information through repeated message-passing layers, coupling the distance over which information travels with the number of nonlinear transformations applied. This coupling can make deep architectures difficul...

📖 Read original article


110. F2STNet: Fair and Federated Spectral-Temporal Modeling for Graph Forecasting ​

Author: Jiayi Zhang, Jinfeng Xu, Hewei Wang, Siyuan Cen, Haidong Huang, Yiyao Zhan, Zheyu Chen, Jinjiang You, Ai Jian, Edith C. H. Ngai
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09082v1 Announce Type: new Abstract: Spatiotemporal prediction on graph-structured data is central to traffic forecasting and environmental monitoring, yet decentralized and heterogeneous data complicate both sequence modeling and collaborative training. We propose F$^2$STNet, a federated...

📖 Read original article


111. RAVEN: Frozen Random Graph Reservoirs with Physics-Informed Interaction Fingerprints for Protein-Ligand Binding Affinity Prediction ​

Author: Qingyang Zou, Jiaye Huang, Hangbo Xie, Jiayue Yin, Youyi Song, Jinfeng Liu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09099v1 Announce Type: new Abstract: Quantitative estimation of protein-ligand binding affinity from three-dimensional complex structures is a fundamental task in structure-based computational chemistry and molecular modeling. Reliable prediction remains challenging because available stru...

📖 Read original article


112. Real Data Closes Synthetic-to-Real Gap in Optical Chemical Structure Recognition ​

Author: Yani Guan, Dengpan Dong, Zi Wei, Shuang Luo, Dan Hannah, Yumin Zhang, Kang Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.09100v1 Announce Type: new Abstract: Millions of chemical structures appear in patents and papers only as drawings, and using that information at scale requires reading the drawings. OCSR appears nearly solved on synthetic images yet remains difficult on real documents: the starting recog...

📖 Read original article


113. A Probabilistic Circuit-Induced Pseudo-Metric for Out-of-Distribution Detection ​

Author: Bhumika K, Vidhya S, Narayanan C Krishnan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09117v1 Announce Type: new Abstract: Probabilistic Circuits (PCs) are tractable generative models whose internal nodes encode a hierarchy of probabilistic sum- maries over different variable scopes. Existing PC-based out- of-distribution (OOD) detection methods ignore this hierar- chy, re...

📖 Read original article


114. MARA: Flow-Matching-Guided Multi-Agent Resource Allocation for Computational Resource Efficient Learning ​

Author: Hanye Zhao, Muning Wen, Yong Yu, Weinan Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09130v1 Announce Type: new Abstract: Allocating limited computation among concurrent learning tasks is difficult when each task must reach a target loss before a deadline but its required training effort is unknown. Existing approaches combine online loss prediction with adaptive resource...

📖 Read original article


115. SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization ​

Author: Gyudong Kim, Wonjun Han, Young Geun Kim
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2608.09160v1 Announce Type: new Abstract: Query-Key Normalization (QK-Norm) improves the training stability and quality of modern Large Language Models (LLMs). However, under Tensor Parallelism (TP), layerwise QK-Norm introduces additional cross-GPU communication because the normalization fact...

📖 Read original article


116. Tabular Numeric Stretch Transformation ​

Author: Zihao Ye, Juyong Kim, Johnna Sundberg, Burak Varici, Pradeep Ravikumar
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09162v1 Announce Type: new Abstract: Tabular data presents unique challenges for deep learning due to its heterogeneous nature, where numeric features exhibit diverse distributions, scales, and statistical properties. Although recent advances have improved how models learn from tabular da...

📖 Read original article


117. A Time-Frequency Dual-Domain Multi-Scale Convolutional Neural Network for Bearing Fault Diagnosis under Strong Noise ​

Author: Yanxi Ding, Tingyue Jia
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09174v1 Announce Type: new Abstract: To address the degradation of bearing fault diagnosis accuracy under strong noise, this paper proposes a time-frequency dual-domain multi-scale convolutional neural network. The time-domain branch employs three parallel convolutional kernels to capture...

📖 Read original article


118. FedA2L: Adaptive layer-wise learning rate adjustment in decentralized federated learning ​

Author: Van Truong Vo, Khoa Nguyen, Taehong Kim
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.DC

arXiv:2608.09208v1 Announce Type: new Abstract: Decentralized intelligence systems with heterogeneous devices and limited coordination increasingly rely on decentralized federated learning (DFL). However, DFL suffers from convergence inefficiency under data heterogeneity due to the use of a uniform ...

📖 Read original article


119. Beyond Solvability: Task Learnability as a Static Prior for LLM RL Post-Training ​

Author: Ting Zhou, Zhenqing Ling, Daoyuan Chen, Qianli Shen, Yilun Huang, Ying Shen, Yaliang Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09217v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a central post-training paradigm for eliciting reasoning capabilities in large language models, yet uniform task sampling allocates compute without regard to differences in how tasks respond to optimization. Exist...

📖 Read original article


120. Online Learning of Scale Parameters in Score-Driven Filters ​

Author: Fabrizio Lillo, Giulia Livieri, Gianluca Palmari
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ME, stat.ML, stat.TH

arXiv:2608.09218v1 Announce Type: new Abstract: Score-driven filters multiply a scaled log-likelihood score by a gain that controls the update magnitude. We treat this gain as a decision variable and study its online learning. Conditional on the current state, observation, score, and scaling rule, e...

📖 Read original article


121. FedTVD: Balancing Data Quality and Quantity for Robust Federated Learning ​

Author: Radwan Selo, Majid Kundroo, Taehong Kim
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.DC

arXiv:2608.09221v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative model training across distributed client devices while preserving data privacy. However, FL faces significant challenges due to data heterogeneity, particularly in terms of label distribution skewness and v...

📖 Read original article


122. Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Distillation ​

Author: Yuki Ichihara, Naoto Iwase, Mohammad Atif Quamar, Junpei Komiyama
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09228v1 Announce Type: new Abstract: On-Policy Self-Distillation (OPSD) is commonly interpreted as the transfer of privileged information: a teacher observes the verified solution to the target problem and supervises the student's trajectory. However, this interpretation conflates two eff...

📖 Read original article


123. DreOPD: Degraded-Reference Extrapolative On-Policy Distillation for Flow-matching Models ​

Author: Mingfeng Lin, Chengfei Cai, Lin Xu, Yuxiang Wei, Liang Han
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.09233v1 Announce Type: new Abstract: Flow-matching models are now a mainstream method to image generation, but its adaptation to diverse downstream scenarios typically relies on post-training, which may cause conflicts among task-specific optimization objectives. Reinforcement learning en...

📖 Read original article


124. Label Granularity Skew in Federated Learning with Hierarchical Image Classification ​

Author: Jaeheon Kim, Hokeun Kim, Bong Jun Choi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09236v1 Announce Type: new Abstract: Federated learning enables privacy-preserving collaboration across distributed devices without centralizing local data. However, clients may differ not only in data distributions but also in domain knowledge and annotation capabilities. In this paper, ...

📖 Read original article


125. Multimodal Federated Learning under Dual-Axis Modality Missingness ​

Author: Adiba Orzikulova, Jaehyun Kwak, Jaemin Shin, Yunqi Guo, Xiaomin Ouyang, Guoliang Xing, Steven Euijong Whang, Sung-Ju Lee
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09240v1 Announce Type: new Abstract: Multimodal federated learning (FL) supports collaborative modeling in privacy-sensitive health-sensing and medical settings, but realistic deployments often exhibit dual-axis modality missingness: clients have different modality sets, and individual sa...

📖 Read original article


126. FEAST: Federated Shared-Space Training for Resource-Heterogeneous Clients ​

Author: Bostan Khan, Masoud Daneshtalab
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2608.09250v1 Announce Type: new Abstract: Federated learning (FL) must serve devices with varying computational capabilities. A fixed model cannot suit all devices, while training one model per deployment limit is costly. Federated supernet training instead learns one elastic model with differ...

📖 Read original article


127. Full-Feature versus Limited-Input Machine Learning for Residential Energy Estimation: A Comparative Analysis of RECS and ResStock Under Realistic Input Constraints ​

Author: Aditya Ramnarayan, Fatih Evren, Patti Gunderson, Samuel Rosenberg
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09255v1 Announce Type: new Abstract: Residential energy estimates are often needed before detailed envelope characteristics, equipment efficiencies, infiltration, sensor, or billing data are available. This study quantifies the trade-off between predictive accuracy and input accessibility...

📖 Read original article


128. SoftmaxGRPO: Learning to Reason using Softmax Advantage Group Estimation ​

Author: Jefferson Hernandez, Jaywon Koo, Zilin Xiao, Chen Wei, Vicente Ordonez
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09271v1 Announce Type: new Abstract: Group-based reinforcement learning objectives such as GRPO can allocate learning signal poorly across prompt difficulty: under binary rewards, group normalization induces a divergent weighting on easy prompts. We introduce Softmax Advantage Group Estim...

📖 Read original article


129. VeinCast: Physics-Guided Dynamic Field Graphs with Graph-Conditioned Fusion for Global Medium-Range Weather Forecasting ​

Author: Zhisheng Chen, Jinhan Li, Yuxuan Li, Yuan Gao, Hao Wu, Zheng Lu, Jinlong Du, Kun Wang, Bo An
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09286v1 Announce Type: new Abstract: Global medium-range weather forecasting requires modeling structured yet state-dependent interactions among heterogeneous atmospheric fields. Existing data-driven models largely learn these interactions implicitly, whereas equation-level physical const...

📖 Read original article


130. Beyond the Capability Boundary: Zeroth-Order Optimization for Self-Evolving LLM Agents ​

Author: Bingzhen Liu, Xiaomeng Fan, Yuwei Wu, Zhi Gao, Mingyang Gao, Chuanhao Li, Yunde Jia
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.09292v1 Announce Type: new Abstract: Self-evolving methods improve the capabilities of LLM agents by sampling trajectories from the underlying LLMs and learning from these trajectories. However, these methods struggle to learn beyond the inherent capability boundary of the agents, since t...

📖 Read original article


131. Targeted Label-Flipping and Oversampling Attacks on Federated Conditional GANs ​

Author: Panav Shah, Avishek Ghosh
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09314v1 Announce Type: new Abstract: In a federated learning setup for GANs, several adversarial attacks are possible. One such attack is label flipping, in which malicious clients deliberately alter label information during local training in order to manipulate the global generator. The ...

📖 Read original article


132. MaxModShift: Model Privacy via Designed Shifts ​

Author: Nomaan A. Kherani, Urbashi Mitra
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, eess.SP, math.IT

arXiv:2608.09328v1 Announce Type: new Abstract: Model learning by an eavesdropper is treated as an estimation problem in a federated environment. The Fisher Information Matrix for the eavesdropper's estimation problem is driven to singularity through a signaling design; this ensures that the eavesdr...

📖 Read original article


133. Hallucinations and Constraints : Regulating surgical workflow recognition beyond accuracy ​

Author: John S. H. Baxter, Pierre Jannin
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09332v1 Announce Type: new Abstract: Hallucinations are a major concern for the integration of artificial intelligence into medicine, although less explored in the realm of medical image processing. Unlike problems in natural text understanding and reasoning therewith, determining whether...

📖 Read original article


134. In-Context Density Estimation for Tabular Data ​

Author: Patryk Marsza{\l}ek, Jacek Tabor, Marek 'Smieja
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09348v1 Announce Type: new Abstract: Density estimation underlies many unsupervised tasks on tabular data such as anomaly detection, out-of-distribution detection, and data augmentation. Although all these problems reduce to questions about where probability mass lies, they are typically ...

📖 Read original article


135. Test-Time Augmentation for LLMs: When Input Diversity Beats Output Diversity at Matched Compute ​

Author: Nikita Kozodoi, Zainab Afolabi, Jack Butler
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.09351v1 Announce Type: new Abstract: Test-time scaling improves LLM accuracy but multiplies inference cost, making the accuracy gained per unit of compute the metric that matters in deployment. Self-consistency is one of the established approaches, which spends this budget entirely on the...

📖 Read original article


136. Beyond Binary: Continuous State Optimization with Graph-Structured Objectives ​

Author: Corinna Cortes, Yishay Mansour, Mehryar Mohri
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09366v1 Announce Type: new Abstract: Large-scale learning systems often face the challenge of balancing multiple, potentially competing objectives, such as fairness, accuracy, and latency. While recent work has formalized this as an optimization problem over binary states, many real-world...

📖 Read original article


137. Imaginative Generative AI: Crossing the Entropy Wall into Worlds Beyond Imitation ​

Author: Farzan Farnia, Hossein Goli, Amin Gohari
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.09385v2 Announce Type: new Abstract: Generative AI models are primarily designed to imitate the data distribution, an objective that neither corrects diversity lost by a learned generator nor defines how generation should extend beyond the diversity of the data itself. We introduce Imagin...

📖 Read original article


138. From Objectives to What Models Learn: A Landau Theory of Invariant Learning ​

Author: Pinli Wang, Yue He, Peng Cui
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09396v1 Announce Type: new Abstract: Invariant learning seeks representations that remain predictive across environments, yet the behavior of its objectives along the regularization path is often opaque. We address this objective-behavior gap by viewing representation learning as multimod...

📖 Read original article


139. Why Post-Norm Transformers Collapse: Attention Amplification and Gradient Repair Failure ​

Author: Xingjian Wang, Qingyu Han, Xiaodong Luo, Yin Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09417v2 Announce Type: new Abstract: Deep decoder-only Transformers often replace the original Post-Norm architecture with Pre-Norm variants because Post-Norm training is highly sensitive to warmup and learning rate under conventional initialization schemes. Although prior work has identi...

📖 Read original article


140. LITEWAY: LIghtweight HAR via Temporal Efficient highWAY ​

Author: Dominique Nshimyimana, Vitor Fortes Rey, Mengxi Liu, Bo Zhou, Paul Lukowicz
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC

arXiv:2608.09421v1 Announce Type: new Abstract: Wearable human activity recognition (HAR) remains challenging due to the computational and energy constraints of deep learning models on resource-limited devices. Existing lightweight approaches often rely on recurrent architectures (e.g., GRU and LSTM...

📖 Read original article


141. How Simple Can It Get? From Interpretable Equations to Readable Rules for Financial Decision Making ​

Author: Adia Lumadjeng, Ilker Birbil, Erman Acar
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09433v1 Announce Type: new Abstract: In regulated domains such as finance, a model that cannot be explained cannot be deployed, yet many interpretable classifiers defeat their own purpose by producing formulas with dozens of features that no regulator could read. We take the reverse direc...

📖 Read original article


142. Depth-adaptive Inference of Looped Language Models via Continuous Depth Batching ​

Author: Kristian Schwethelm, Daniel Rueckert, Georgios Kaissis
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.DC

arXiv:2608.09444v1 Announce Type: new Abstract: A main promise of looped language models (LMs) is depth-adaptive inference. By iterating a block of shared layers a variable number of times, the model can use less compute for "easy" tokens and more for "hard" ones. However, this adaptivity breaks sta...

📖 Read original article


143. WDL-OPD: Weak-Driven On-Policy Distillation via Mixture-Constrained Co-Training ​

Author: Zehao Chen, Gongxun Li, Tianxiang Ai, Yifei Li, Zixuan Huang, Wang Zhou, Tao Huang, Fuzhen Zhuang, Xianglong Liu, Jianxin Li, Deqing Wang, Yikun Ban
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09447v1 Announce Type: new Abstract: On-policy distillation (OPD) aligns a student with a teacher on trajectories sampled from the student itself, reducing the train-test state mismatch of offline distillation. The same feedback loop can nevertheless be unstable: each update changes both ...

📖 Read original article


144. From Approachability Residuals to Anytime-Valid Evidence: The Online Convex Geometry of Testing by Betting ​

Author: Jinze Zhao
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09450v1 Announce Type: new Abstract: Betting-based sequential tests and Blackwell approachability are linked by a rate-explicit reduction through support-function residuals. For a compact convex target $S$ and vector observations $r_t$, an OCO learner selects a predictable normal $w_t$ an...

📖 Read original article


145. Learning to Modulate, Not to Cycle: Soft Actor---Critic Recovers Inverter-Style Heat-Pump Control ​

Author: Faizan Ahmed, Aniket Dixit, James Brusey
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09453v1 Announce Type: new Abstract: On--off cycling is the main cause of compressor wear in residential heat pumps, yet reinforcement learning (RL) controllers for buildings typically optimise only energy cost and thermal comfort, ignoring how much the learned policy cycles. We add a lev...

📖 Read original article


146. Flow-based conditional cardiac anatomy generation for virtual cohorts ​

Author: Konstantinos Kevopoulos, Beatrice Moscoloni, Benjamin Alheit, Cameron Beeche, Julio A. Chirinos, Alexander Heinlein, Mathias Peirlinck
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, q-bio.QM, q-bio.TO

arXiv:2608.09460v1 Announce Type: new Abstract: Cardiac digital twin research is moving from subject-specific anatomical replicas toward virtual cohorts that represent clinically relevant population subgroups. Yet access to representative imaging-derived anatomy datasets remains limited by cohort si...

📖 Read original article


147. MixFormer: Linear Transformer with Mixture of Memory Experts ​

Author: Yu Guo, Lei Duan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09468v1 Announce Type: new Abstract: State Space Models (SSMs), as a mainstream research direction of linear Transformers, aim to achieve higher efficiency than standard Transformers in long-context modeling. However, existing SSMs suffer from limited input adaptivity and constrained memo...

📖 Read original article


148. Hierarchical rank-evolving representation for physics-informed neural networks ​

Author: Ruoyang Su, Xi-Le Zhao, Kun Li, Liang Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, physics.comp-ph

arXiv:2608.09483v1 Announce Type: new Abstract: Recently, tensor-based physics-informed neural networks (T-PINNs) have received increasing attention. However, existing T-PINNs still face a fundamental challenge: they mainly rely on pre-specified low-rank tensor decompositions with manually tuned ran...

📖 Read original article


149. When Do Task Vectors Interfere? Mapping the Validity Boundaries of Weight-Space Composition ​

Author: Chencheng Zhu, Xiaoyang Li, Taotao Cai
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09490v1 Announce Type: new Abstract: Task arithmetic treats fine-tuning displacements as composable directions in weight space, yet it remains unclear when parameter addition reflects predictable changes in model function. We separate parameter geometry from functional geometry and measur...

📖 Read original article


150. Tracking the Best Strategy in an Extensive-Form Game ​

Author: Stephen Pasteris, Rahul Savani, Theodore Turocy
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09501v1 Announce Type: new Abstract: We consider the extensive-form bandit problem where on each trial the learner plays an extensive-form game against an oblivious adversary. We focus on the notion of switching regret, which measures the expected performance of the learner against that o...

📖 Read original article


151. Generalized Convexity and Smoothness via Conjugate Duality: Optimization Theory for Deep Neural Networks ​

Author: Binchuan Qi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09523v1 Announce Type: new Abstract: Deep neural network (DNN) training with stochastic gradient descent (SGD) and its variants achieves strong empirical performance, yet classical optimization theory does not fully explain this success. This limitation arises because conventional analyse...

📖 Read original article


152. Dual-Adversarial Safety Alignment: Cultivating Intrinsic Threat Comprehension in LRMs ​

Author: Hongli Shen, Shaopeng Fu, Qinbo Zhang, Jian Li, Di Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2608.09542v1 Announce Type: new Abstract: Large reasoning models (LRMs) achieve remarkable success on complex tasks but remain vulnerable to harmful prompts that induce unsafe outputs. Recent methods align LRMs using direct refusals or safety rationales, yet often focus on prompt patterns rath...

📖 Read original article


153. Training-Free Universal Approximation by Prompting Random Transformers ​

Author: Alexander Hsu, Rongjie Lai
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, math.ST, stat.TH

arXiv:2608.09558v1 Announce Type: new Abstract: How expressive is prompting a transformer? Answering this question is important for separating the roles of prompting, architecture, and pretraining in transformer models, and for determining whether task-specific behavior must be stored in model weigh...

📖 Read original article


154. Hyperbolic Multimodal Continual Learning ​

Author: Jiahong Liu, Ming Shen, Xiaohao Liu, Rex Ying, Menglin Yang, Tat-Seng Chua, Irwin King
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09572v1 Announce Type: new Abstract: Hyperbolic geometry has recently emerged as a powerful representation space for multimodal learning, as it naturally captures hierarchical semantic structure across modalities. Despite this progress, how such representations behave under continual lear...

📖 Read original article


155. LEED: Local Embedding Evolution Distance for over-smoothing estimation and virtual node selection in GNN ​

Author: Killian Cressant, Pedro B. Velloso
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09596v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) suffer from two fundamental limitations: over-smoothing, where node representations become indistinguishable with depth, and over-squashing, where long-range information is compressed through limited message-passing channel...

📖 Read original article


156. Bayesian Symbolic Regression with Entropic Reinforcement Learning ​

Author: Oussama Boussif, Mohammed Mahfoud, Younesse Kaddar, Moksh Jain, Sida Li, Damiano Fornasiere, Xiaoyin Chen, Yoshua Bengio, Esmeralda S. Whitammer
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09617v2 Announce Type: new Abstract: Symbolic regression is the problem of finding an algebraic expression describing a stochastic dependence of a target variable on a set of inputs. Unlike forms of regression that fit parameters assuming a fixed model structure, symbolic regression is a ...

📖 Read original article


157. Satellite Trajectory Optimization via Proximal Policy Optimization for Space Debris Avoidance ​

Author: Logan Luna (Georgia Institute of Technology), Juan Ortiz Couder (Embry-Riddle Aeronautical University), Raul Alejandro Vargas-Acosta (Embry-Riddle Aeronautical University)
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.09628v1 Announce Type: new Abstract: Collision avoidance systems are commonly used to avoid fragmentation events occurring in Low-Earth Orbit (LEO) and Geosynchronous Equatorial Orbit (GEO). However, these events have been growing in frequency as orbital congestion worsens with the launch...

📖 Read original article


158. Deep Learning Imputation of Missing Radius of Maximum Winds (Rmax) Values in Tropical Cyclone Best-Track Data ​

Author: Swastik Agrawal, Nishkal Hundia, Ziyue Liu, Michelle Bensi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2608.09683v1 Announce Type: new Abstract: Probabilistic coastal hazard assessments require accurate characterization of tropical cyclone (TC) parameters, yet datasets often contain missing records for the radius of maximum winds (Rmax), a key variable in Joint Probability Method analyses. This...

📖 Read original article


159. FedOrbit: Adaptive Personalized Federated Learning for Non-IID LEO Satellite Constellations ​

Author: Satwat Bashir, Tasos Dagiuklas, Muddesar Iqbal
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09687v1 Announce Type: new Abstract: Federated learning (FL) in Low Earth Orbit (LEO) satellite constellations is affected by non-IID data and irregular ground-station visibility, both driven by orbital geometry. Global aggregation performs poorly when orbit-level class distributions are ...

📖 Read original article


160. Confusion-Geometry Rebalancing for Long-Tailed Adversarial Training ​

Author: Mengnan Zhao, Geyong Min, Lihe Zhang, Tianhang Zheng, Jie Cui
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09688v1 Announce Type: new Abstract: Adversarial training under long tailed distributions suffers from a dual imbalance: the class imbalance skews the training objective toward head classes, and the adversarial inner maximization may further amplify this bias. Existing methods mitigate th...

📖 Read original article


161. Recurrent Neural Networks Beyond Time: Learning from Multiple Ordered Projections ​

Author: Vagan Terziyan, Artur Terziian, Oleksandra Vitko
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09690v1 Announce Type: new Abstract: Recurrent neural networks (RNNs) are widely used for sequence learning, yet their application is commonly associated with temporal data, although recurrent computation fundamentally operates on ordered sequences rather than on time itself. Building on ...

📖 Read original article


162. Evaluating Generative Time-Series Models on Data with Point Masses ​

Author: Jian Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09692v1 Announce Type: new Abstract: Many of the series that generative time-series models are benchmarked on place a large probability mass on a single value --- it does not rain, no ride is requested, no part is ordered. We report what happens when such data is evaluated carefully. Firs...

📖 Read original article


163. PET/CT Radiogenomic Mutation Prediction in Non-Small Cell Lung Cancer Using Multi-Label Learning ​

Author: Mona Furukawa, Sai Hyne, Daniel R. McGowan, Bart{\l}omiej W. Papie.z
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09721v1 Announce Type: new Abstract: Lung cancer remains one of the leading causes of cancer- related mortality worldwide. Although targeted therapies have improved outcomes for patients with non-small cell lung cancer (NSCLC), they rely on mutation profiling through tissue biopsy, an inv...

📖 Read original article


164. Rethinking Factor Sharing in Federated LoRA: A Rank-Aware Adaptive Approach ​

Author: Xinyi Xu, Bingnan Xiao, Shuang Qin, Gang Feng, Tony Q. S. Quek
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC

arXiv:2608.09742v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) represents large language model (LLM) updates with two compact matrix factors, i.e., $A$ and $B$, providing an efficient way to fine-tune large models in federated learning paradigm. Inspired by the asymmetric roles of the Lo...

📖 Read original article


165. SR-OPSD: Self-Referenced On-Policy Self-Distillation ​

Author: Zhuo Sun, Entong Li, Yanlong Zhao, Xiaoyuan Cheng, Wenxuan Yuan, Kaiyu Li, Che Liu, Huihang Liu, Harrison Bo Hua Zhu, Li Zeng
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.09745v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) converts feedback into dense token-level supervision on trajectories generated by the policy to be optimized, providing a useful complement to reinforcement learning with sparse outcome rewards. However, the self-teac...

📖 Read original article


166. MoNo: Multiscale Optimal Transport Neural Operator for Solving PDEs on General Geometries ​

Author: Zijiang Yang, Xiaomeng Wu, Dongmei Fu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09764v1 Announce Type: new Abstract: Transformer-based neural operators have achieved substantial progress in solving Partial Differential Equations (PDEs) by projecting spatial observations into compact latent tokens and learning physical interactions in latent spaces. However, we reveal...

📖 Read original article


167. ReliableNet: A Chance-Constrained Approach to Trustworthy Classification in Deep Learning ​

Author: Ange-Cl'ement Akazan, Ineza Remy Mugenga, Abebe Geletu, Jean Medard Ngnotchouye, Issa Karambal
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09768v1 Announce Type: new Abstract: A prediction that is both confident and wrong is a critical reliability failure because it can bypass abstention and human review precisely when the model is mistaken. Empirical risk minimization (ERM) controls average loss but not this failure directl...

📖 Read original article


168. Parameter Exploration for RLVR via Variational Learning ​

Author: Vatsal Venkatkrishna, Nico Daheim, Iryna Gurevych
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.09805v1 Announce Type: new Abstract: Exploration has been a focus of reinforcement learning research for a long time. Recently, there has been growing evidence that it is also an important ingredient in LLM reinforcement learning recipes that can significantly impact downstream performanc...

📖 Read original article


169. Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA ​

Author: Mind Lab, :, Vin Bo, Asher Cai, Jingwei Cao, Song Cao, Vic Cao, Amelia Chen, Andrew Chen, Kaijie Chen, Cleon Cheng, Steven Chiang, Kaixuan Fan, Hera Feng, Huan Feng, Arthur Fu, Jun Gao, Pyke Han, Nolan Ho, Ori Hong, Hailee Hou, Piers Hua, Charles Huang, Miles Jiang, Nora Jiang, Yuyi Jiang, Qiuyu Jin, Fancy Kong, Kuss Koo, Jaron Lee, Andrew Lei, Alexy Li, Dawn Li, Lucian Li, Ray Li, Ricardo Li, Smith Li, Theo Li, Allen Lin, Elliot Lin, Fan Lin, Chen Ling, Kairus Liu, Kieran Liu, Logan Liu, Neo Liu, Xiang Liu, Yuxin Lu, Maeve Luo, Pony Ma, Verity Niu, Cole Qiao, Guian Qiu, Vince Qu, Sentry, Niko Song, Vincent Wang, Bo Wu, Rio Yang, Evelyn Ye, Fiona Ye, Ina Ye, Regis Ye, Josh Ying, Atlas Zeng, Danney Zeng, Salmon Zhan, Anya Zhang, Di Zhang, Mia Zhang, Sueky Zhang, Wei Zhao, Ada Zhou, Adrian Zhou, Yuhua Zhou, Juno Zhu, Murphy Zhuang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.09819v1 Announce Type: new Abstract: Macaron-V1 is an open agent-model family for experiential intelligence: learning from experience in real environments and continuing to learn after deployment. It is organized around two system goals. Adaptation is pursued through recursive improvement...

📖 Read original article


170. Distill Skills into Weights, Not Prompts: Abstract Skills as Privileged Signals for On-Policy Self-Distillation ​

Author: Yubo Jiang, Fengying Xie, Zhiguo Jiang, Haopeng Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.09826v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards yields no group-relative signal when rollout groups are uniformly correct or uniformly wrong, which account for 63.0-68.0% of groups in our experiments. We propose SKALD (Skill-Anchored Latent Distillation...

📖 Read original article


171. Multi-Agent AI Safety as an Institutional Design Problem ​

Author: Abdullah X
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA

arXiv:2608.09828v1 Announce Type: new Abstract: AI agents increasingly work inside systems that govern how they delegate tasks, move information, execute actions, and use shared resources. Recent work already shows that deployment rules can change collective behavior. Here we ask which parts of an A...

📖 Read original article


172. Deep Multimodal Wearable Sensor Fusion for Detection of Body-Focused Repetitive Behaviors ​

Author: Samaneh Rezaeimanesh, Mohsen Behradfar, Mohammad Fili, Guiping Hu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09830v1 Announce Type: new Abstract: Body-focused repetitive behaviors, such as hair pulling and skin picking, are compulsive motor actions commonly associated with obsessive-compulsive and anxiety disorders. Their early, objective detection remains difficult because the movements are sub...

📖 Read original article


173. Real-Time Climate Risk Assessment for Supply Chain Resilience: A Data-Driven Nowcasting Framework for Colombian Agriculture ​

Author: Hernan J. Silva-Sosa
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09846v1 Announce Type: new Abstract: This paper presents a methodological framework for real-time climate risk assessment using data-driven nowcasting techniques to enhance supply chain resilience in Colombian agricultural contexts. Climate variability in Colombia, characterized by irregu...

📖 Read original article


Author: Valentijn Oldenburg, Floris de Kam, Stef de Wildt, Jarno Nilson Balk
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.SI

arXiv:2608.09899v1 Announce Type: new Abstract: In fair ranked link prediction, demographic parity ($\Delta_\mathrm{DP}$) is a common fairness metric. Yet, Mattos et al. (2025) argue that it fails to detect exposure bias because it ignores where links appear in the ranking. In this study, we reprodu...

📖 Read original article


175. Emotion in an active inference model of human driving ​

Author: Julian F. Schumann, Johan Engstr"om, Ran Wei, Jens Kober, Martijn Wisse, Arkady Zgonnikov
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.HC, cs.LG, cs.RO

arXiv:2608.07480v1 Announce Type: cross Abstract: Active inference has emerged as a principled framework for modeling adaptive behavior by balancing goal-directed action with uncertainty reduction. It has been successfully applied across biological and artificial systems, including recent work on hu...

📖 Read original article


176. Knowing You Is Everything: LLM Agents Achieve Near-Perfect Profile-Consistent Reaction Prediction in Social Media Simulation ​

Author: Ljubisa Bojic, Ljiljana Matic, Joerg Matthes, Milan Cabarkapa, Bojana Dinic, Jue Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.LG, cs.MA

arXiv:2608.07498v1 Announce Type: cross Abstract: Autonomous AI agents in social media present concrete risks to democratic discourse and platform governance, while also offering tools for pre-deployment recommender system testing. A central open question is whether persona-prompted LLMs can simulat...

📖 Read original article


177. Explainable Machine Learning in Healthcare: Methods, Interpretation, and Applications for Clinical Research ​

Author: Krishna Padmanabhan, Minxin Lu, Dai Feng, Natalia KanDobrosky, Sai Konduri, Heather J. Litman, Achilleas Livieratos
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CY, cs.LG

arXiv:2608.07522v1 Announce Type: cross Abstract: We present a structured review of commonly used Explainable machine learning (XML) methodologies, including global and local interpretability tools such as SHapley Additive exPlanations (SHAP), Local Interpretable Model-Agnostic Explanations (LIME), ...

📖 Read original article


178. Training Variable Long Sequences with Data-Centric Parallel ​

Author: Geng Zhang, Xuanlei Zhao, Kai Wang, Yang You
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.07524v1 Announce Type: cross Abstract: Training deep learning models on variable long sequences poses significant computational challenges. Existing methods force a difficult trade-off between efficiency and ease-of-use. Simple approaches use static configurations that cause workload imba...

📖 Read original article


179. The Knowing-Saying Gap: When Probes See Errors that Confidence Misses ​

Author: Jyotin Goel, Ipshita Bandyopadhyay, Justin Shenk
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.07528v1 Announce Type: cross Abstract: Linear probes detect corrupted context in language models with near-perfect accuracy, yet this does not translate into reliable failure prediction. The result is a dissociation with direct implications for deployment monitoring. Across multi-hop arit...

📖 Read original article


180. Dynamic Coalition Formation and Communication Pricing in Skill-Based Agentic AI Systems ​

Author: Mojtaba Eslami
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, econ.TH

arXiv:2608.07532v1 Announce Type: cross Abstract: Modern agentic AI systems combine multiple large language model agents with heterogeneous skills, yet most architectures either fix communication in advance or allow full broadcast. Both can be inefficient because token cost, latency, redundancy, and...

📖 Read original article


181. An AI Scientist that Doesn't Drift: Taste, Structure, and Falsifiable Findings in a Quadruped Navigation Research Loop ​

Author: Yiwen Zhang, Eloise Zeng, Jaeha Lee, Tony Yue Yu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA, cs.RO

arXiv:2608.07542v1 Announce Type: cross Abstract: Autonomous research loops driven by large language models can run machine-learning experiments at scale but tend to drift toward local refinements of whichever metric they optimise rather than testing the hypotheses that motivate the experiments. We ...

📖 Read original article


182. DarwinX: Evolving Agent Harnesses Through Natural Selection ​

Author: Yifan Zhang, Yutong Dai, Juntao Tan, Luyu Yang, Rishi Mullur, Thai Hoang, Zhiyuan Hu, James Zhu, Phil Mui, Silvio Savarese, Ran Xu, Zeyuan Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG, cs.SE

arXiv:2608.07545v1 Announce Type: cross Abstract: An LLM agent's capability depends not only on model weights but on its harness: prompts, tools, skills, and control flow. Self-improvement loops already edit harnesses, yet single-lineage search is path-dependent and local wins often regress other ta...

📖 Read original article


183. Auditing Medical Vision-Language Models on Chest Radiographs: Estimating Reference Agreement Across Institutions ​

Author: Pengyang Yu, Yiou Wang, Zhongping Dong, Sahraoui Dhelim, Chun-Mei Feng, M. Tahar Kechadi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.07550v1 Announce Type: cross Abstract: Vision-language models return structured chest-radiograph findings through interfaces exposing no confidence score, so a receiving institution cannot read off how far to trust an individual judgment. Whether agreement with an institution's reference ...

📖 Read original article


184. MVMD: A Multi-View Approach for Enhanced Mirror Detection ​

Author: Yidan Shen, Yu Wen, Chen Zhang, Xin Fu, Renjie Hu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.07559v1 Announce Type: cross Abstract: In 3D reconstruction, mirrors introduce significant challenges by creating distorted and fragmented spaces, resulting in inaccurate and unreliable 3D models. As 3D reconstruction typically relies on multi-view images to capture different perspectives...

📖 Read original article


185. Mechanistic Interpretability-Guided Selective Fine-Tuning of Vision-Language Models for Centimeter-Level Flood Depth Estimation ​

Author: Nafis Fuad, Xiaodong Qian, Dongxiao Zhu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.07562v1 Announce Type: cross Abstract: Urban flooding poses an escalating threat to transportation infrastructure, yet no operational system provides real-time, street-level flood-depth estimates at centimeter resolution. This paper presents three vision-language models fine-tuned for con...

📖 Read original article


186. Latent-Frequency Validity: Fast Spectral Editing with Screened Video-VAE Transfer Operators ​

Author: Bowen Xue, Jiafeng Xiong, Xin Quan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.07569v1 Announce Type: cross Abstract: Direct spectral editing in video-VAE latents can control noise, flicker, smoothness, and frequency content without a decode--filter--reencode pass. However, video VAEs may redistribute pixel-space frequency bands across latent channels, and latent ed...

📖 Read original article


187. Beyond Isotropic Assumptions: Continuity-Constrained Segmentation and GPU Morphometry for Nanoscale GBM Analysis ​

Author: Arash Fatehi, Robin Ebbestad, Linus Butt, Hans Blom, Sigrid Lundberg, Hannes Olauson, Hjalmar Brismar, David Unnersj"o-Jess, Thomas Benzing, Katarzyna Bozek
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.07575v1 Announce Type: cross Abstract: Confocal microscopy of optically cleared and swelled tissue resolves complex biological structures in 3D, but such acquisitions are highly anisotropic: along the under-sampled axial direction the structure can appear discontinuous, hampering reconstr...

📖 Read original article


188. Real-time physics inversion for retrieval of sub-pixel wildfire temperatures from VSWIR imaging spectroscopy ​

Author: William R. Keely, Philip G. Brodrick, Katherine Mistick, Adam Chlus, Robert O. Green, Philip E. Dennison
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, astro-ph.IM, cs.LG

arXiv:2608.07580v1 Announce Type: cross Abstract: In this work, we present a wildfire temperature retrieval framework for VSWIR imaging spectroscopy data, employed on data from NASA's Airborne Visible Infrared Imaging Spectrometer (AVIRIS-3). The retrieval framework utilizes a full-physics approach ...

📖 Read original article


189. RouteGuard: Certifying Routing Gain in LLM Multi-Agent Systems When Complementarity Is Not Enough ​

Author: Anchen Sun, Kaiqi Yang
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.07583v1 Announce Type: cross Abstract: Multi-agent LLM systems route among model-backed advisors, yet a deployer rarely knows before shipping whether routing will help at all. Prevailing routers optimize a gate's AUC and presume that advisor complementarity suffices. We show that neither ...

📖 Read original article


190. LAVE: Latent Visual Evidence-Enhanced Planning for Video Tool-use Agents ​

Author: Zijian Wang, Junnan Zhu, Rongzhen Li, Xiao Liu, Guohui Xiang, Quan Lu, Lijia Liu, Yining Wang, Jiang Zhong, Kaiwen Wei
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.MA

arXiv:2608.07585v1 Announce Type: cross Abstract: Long-video understanding requires models to efficiently acquire and reuse sparse visual evidence from long and redundant video streams. Recent video tool-use agents address this challenge by iteratively invoking visual Tools at different temporal sca...

📖 Read original article


191. MAGIC-SSCIL: Manifold Anchoring and Geometric Incremental Calibration for Semi-Supervised Class Incremental Learning ​

Author: Yousef Abdi, Mohammad Asadpour, Yousef Seyfari
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.07586v1 Announce Type: cross Abstract: Semi-supervised Class Incremental Learning (SSCIL) is a severe challenge for neural networks, and it is hardest in the exemplar-free setting where no past data may be stored. Existing methods forget catastrophically due to feature drift, and their ps...

📖 Read original article


192. Distribution-Free Conformal Prediction for Steel Fatigue Strength: Marginal Validity Is Not Enough ​

Author: Irene Boruah
Published: 8/11/2026, 4:00:00 AM
Categories: stat.AP, cs.LG

arXiv:2608.07589v1 Announce Type: cross Abstract: Predicting fatigue failure in steel components experimentally is costly because it requires testing across multiple compositions and processing conditions. This has spurred research on data-driven prediction models. Studies using the NIMS MatNavi ste...

📖 Read original article


193. Exploiting chemical shift variability enables recovery of overlapping metabolites from 1H nuclear magnetic resonance spectra ​

Author: Jesper L{\o}ve Hinrich, Pia Susan Mayer, Bekzod Khakimov, S{\o}ren Balling Engelsen, Morten M{\o}rup
Published: 8/11/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG

arXiv:2608.07610v1 Announce Type: cross Abstract: Overlapping peaks and sample-dependent chemical shift variability prevent reliable metabolite recovery from complex biological spectra. This problem is critical in one-dimensional proton (1D 1H) NMR which has become the standard method providing fast...

📖 Read original article


194. Stochastic gradient descent with discontinuity across a manifold ​

Author: Vivek S. Borkar
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR

arXiv:2608.07618v1 Announce Type: cross Abstract: Stochastic gradient descent for a loss function discontinuous across lower dimensional manifolds is analyzed by studying its differential equation limit.

📖 Read original article


195. Controlled Memory Interference in Continual LLM Agents ​

Author: Ao Ding, Hongzong LI, Shiqin Tang, Li Zhang, Liang Chen, Xuyang Chen, Zi Liang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.IR, cs.LG

arXiv:2608.07622v1 Announce Type: cross Abstract: Long-term memory enables AI agents to maintain continuity across sessions, personalize behavior, and evolve through accumulated experience. Yet memory evolution is not simply a process of storing more information: new experiences may reinforce, revis...

📖 Read original article


196. From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Orchestration Framework for Mission-Critical Hospital Information Management Systems ​

Author: Manideep Dhar, Ritwik Singh, Sharat Chandra Kumar Manikonda
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2608.07627v1 Announce Type: cross Abstract: Hospitals are racing to embed AI, while coping with the surge in adaptation of the technology in other industries, into the triage management, documentation, scheduling, and revenue-cycle workflows, yet most deployments remain as fragmented pilots th...

📖 Read original article


197. Readout-Rank Laws for Isotropic Quantum Tangents ​

Author: Marwan Ait Haddou
Published: 8/11/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.07628v1 Announce Type: cross Abstract: Deep parameterized quantum circuits may remain sensitive to a parameter change while the observables retained by a learning model barely respond. We study this separation for a fixed computational-basis measurement. For a pure-state tangent, we compa...

📖 Read original article


198. Embedding Initialization for Unseen Low-resource Languages in Multilingual NMT: A Case Study on Limbum-English Translation ​

Author: Samiratu Ntohsi, Neza David Tuyishimire, Anesu Kafesu, Marvin Ogore, Samuel Oluwajunwonlo Babalola, Oche Ankeli
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.07629v1 Announce Type: cross Abstract: Multilingual neural machine translation models such as NLLB-200 cover 200 languages but leave thousands unsupported, including most Grassfields Bantu languages of Cameroon. When fine-tuning these models for an unseen language, practitioners must choo...

📖 Read original article


199. Contextual Value Alignment via Multilayer Combinatorial Fusion ​

Author: Yuanhong Wu, Djallel Bouneffouf, D. Frank Hsu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2608.07642v1 Announce Type: cross Abstract: Aligning large language models (LLMs) with human values remains a major challenge, especially for trustworthy AI. While existing approaches such as RLHF, CAI, and their variants have achieved promising results, they often rely on a single-agent frame...

📖 Read original article


200. Data collection from highways: a geometric, class-agnostic approach to embedded vehicle counting ​

Author: Lucas Gouveia Omena Lopes, William W. M. Lira, Alexandre M. Lima, Thales M. A. Vieira
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.07643v1 Announce Type: cross Abstract: Traffic data collection is dominated today by deep object detectors followed by tracking-by-detection, a pipeline that presupposes what is often missing in practice: a detector already trained on the class one wants to count. We revisit a purely geom...

📖 Read original article


201. Mendel G\"odel Machine: Recursive Self-Improving Coding Agents via Comparative Evolution ​

Author: Changzhi Liu, Yilun Liu, Sikuan Yan, Volker Tresp, Yunpu Ma
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.07645v1 Announce Type: cross Abstract: Self-improving coding agents that iteratively rewrite their own source code have demonstrated impressive performance on coding tasks. However, existing solutions generally derive self-modification from a single failure trajectory at a time, overlooki...

📖 Read original article


202. Leveraging generative models to assist Monte Carlo sampling ​

Author: Marylou Gabri'e
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.dis-nn, cond-mat.stat-mech, cs.LG, physics.comp-ph

arXiv:2608.07648v1 Announce Type: cross Abstract: Sampling high-dimensional probability distributions is a central task in scientific computing, with applications ranging from Bayesian inference to statistical physics and molecular simulation. Despite decades of methodological developments, two majo...

📖 Read original article


Author: Sana Tonekaboni, Lena Stempfle, Sasha Ronaghi, Corinna Coupette, I. Glenn Cohen, Emily Alsentzer, Marzyeh Ghassemi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.07705v1 Announce Type: cross Abstract: Clinical foundation models trained on large-scale patient data are increasingly used for decision support, screening, and public health. As deployment expands, privacy risk increasingly arises from model-mediated leakage, yet its prevalence and sever...

📖 Read original article


204. Tokenizer Generator Coupling in Medical Image Generation ​

Author: Liam Chalcroft
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.07713v1 Announce Type: cross Abstract: Latent medical image generators usually treat the tokenizer as fixed preprocessing. We test whether this separation is valid in a controlled ChestMNIST study at 64x64, crossing discrete tokenizers, generator families, and sampler settings under a sha...

📖 Read original article


205. LGNNIC: Acceleration of Large-Scale GNN Training using SmartNICs ​

Author: Liad Gerstman, Aditya Dhakal, Dejan Milojicic, Avi Mendelson
Published: 8/11/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.AR, cs.LG, cs.PF

arXiv:2608.07733v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are widely used across domains such as natural sciences, social network analysis, chip design, and recommendation systems. However, as graph sizes grow, storing and processing them entirely on a single-node CPU-GPU system...

📖 Read original article


206. LoRSA: Toward Generalizable Parameter-Efficient Fine-Tuning for Biomedical Downstream Tasks ​

Author: Saed Moradi, Benyamin Ghojogh, M. Hadi Sepanj, Yimin Yang, Ashirbani Saha
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.07749v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning enables the adaptation of vision foundation models to biomedical tasks under limited computational resources, but a single low-rank update can constrain all task-specific changes to one narrow parameter subspace. This ...

📖 Read original article


207. CoCoNav: Conformal Control for Safe Robot Navigation in Crowds ​

Author: Cheng Guo, Mingzhe Ni, Zheng Liang, Yihu Ling, Yuan Hu, Michele Caprio, Daniele Pucci, Wei Pan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.07751v1 Announce Type: cross Abstract: Safe and efficient robot navigation in crowds requires anticipating pedestrian motion despite uncertain and potentially shifting prediction errors. Existing reactive methods can produce oscillatory behavior, while predictive planners often treat fore...

📖 Read original article


208. Who Built This Model? Tracing LLM Lineage via Spectral Fingerprints in Weight Space ​

Author: Yiwei Chen, Bingqi Shang, Sijia Liu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.07786v1 Announce Type: cross Abstract: Open-weight large language models (LLMs) are increasingly developed through complex, multi-stage pipelines, leading to intricate lineage relationships that reflect model origin, ownership, and evolution. Understanding these relationships is important...

📖 Read original article


209. Conformal Calibration for Multi-Modal Regression with Missing Modalities ​

Author: Ilia Azizi
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.07795v1 Announce Type: cross Abstract: Prediction intervals for multi-modal regression with tabular variables, text, images, or other input sources are difficult to calibrate when those sources disagree or one is missing. A single global quantile averages these regimes together instead of...

📖 Read original article


210. Integrating spectral and morphological plant features with decision-tree models for early-season cotton biomass and nitrogen status estimation from multi-year UAV data ​

Author: Vaishali Swaminathan, Nithya Rajan, J Alex Thomasson, Amrit Shrestha, Karem Meza Capcha, Robert Hardin, Pramod Pokhrel
Published: 8/11/2026, 4:00:00 AM
Categories: eess.IV, cs.LG

arXiv:2608.07801v1 Announce Type: cross Abstract: Precision nitrogen (N) management (PNM) for cotton requires in-season monitoring of crop growth parameters and N status indicators to decide fertilizer timing, placement, and application rates for optimal canopy development and yield. This study deve...

📖 Read original article


211. Preserving Item Semantics for Free: Rethinking Token Initialization in LLM-Based Generative Recommendation ​

Author: Donald Loveland, Liam Collins, Bhuvesh Kumar, Danai Koutra, Neil Shah
Published: 8/11/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.07816v1 Announce Type: cross Abstract: Recent advances in generative recommendation (GR) leverage large language models (LLMs) as recommender backbones, enabling LLMs to directly generate recommendations conditioned on item-interaction histories. In these systems, items are often represen...

📖 Read original article


212. Classical $\mathrm{SU}(2)$ Models Match or Exceed Shallow Variational Quantum Circuits on Vision Benchmarks ​

Author: Christopher Fulton, Irene Tsapara, Lawrence Fulton
Published: 8/11/2026, 4:00:00 AM
Categories: cs.PF, cs.LG

arXiv:2608.07822v1 Announce Type: cross Abstract: Quaternion-valued neural networks and variational quantum circuits (VQCs) both derive local transformations from $\mathrm{SU}(2)$ geometry, yet their performance on classical supervised learning remains poorly understood. We compare real-valued, quat...

📖 Read original article


213. Crowd-Sourced Geographies of Income: Using Google Maps Points of Interest as High-Frequency Proxies for Sub-Municipal Income Estimation in Sao Paulo, Brazil ​

Author: Adrienne C. Kinney, Anya Workman, Ademar Takeo Akabane, Jenna Barac, Paulo Fernando Braga Carvalho, Jeova Farias, Fernando Nascimento, Paulo Ricardo da Silva Oliveira
Published: 8/11/2026, 4:00:00 AM
Categories: stat.AP, cs.CY, cs.LG

arXiv:2608.07871v1 Announce Type: cross Abstract: Accurate, up-to-date income data at the sub-municipal scale is essential for social policy in middle-income countries, yet in Brazil it depends on a costly decennial census whose intercensal gap recently exceeded a decade. We test whether the composi...

📖 Read original article


214. GRACE: LLM-Grounded Semantic Metric Spaces for Scalable Mixed-Data Clustering ​

Author: Zihua Yang, Zhencheng Xie, Junyang Chen, Liang Xie, Yiqun Zhang, Mengke Li, Yang Lu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.IT, cs.LG, math.IT

arXiv:2608.07881v1 Announce Type: cross Abstract: Clustering mixed tabular data requires a unified metric space to bridge the inherent heterogeneity between continuous numerical measurements and discrete categorical symbols. Traditionally, algorithms rely entirely on dataset-internal statistics to e...

📖 Read original article


215. Auditing Instruction-Trajectory Mismatches in Multimodal Robot Demonstrations ​

Author: Simon Holk, Ryosuke Takanami, Tatsuya Matsushima, Yusuke Iwasawa, Yutaka Matsuo, Yueh-Hua Wu, Kei Ota
Published: 8/11/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.07895v1 Announce Type: cross Abstract: Robot demonstration datasets used to train vision-language-action policies can contain a subtle but harmful failure mode: trajectories that are behaviorally correct but paired with the wrong language instruction. We study post-hoc auditing of these I...

📖 Read original article


216. The Spectral Neuron ​

Author: Alex Shtoff
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.08003v1 Announce Type: cross Abstract: As machine learned models increase in complexity and expressive power, features of simpler models, such as interpretability and control over the shape of the modeled function are lost. On the one edge of the spectrum we have simple linear models are ...

📖 Read original article


217. BASIS: Breach-Aware Selective Prompt Injection Shielding with Prefill Attention Probes ​

Author: Laiqiao Qin, Tianqing Zhu, Longxiang Gao, Wanlei Zhou
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.08027v1 Announce Type: cross Abstract: Prompt injection is a critical security threat in large language model (LLM) applications, where attackers hijack model behavior by embedding malicious instructions in user or external data. Existing detection methods only detect the presence of inje...

📖 Read original article


218. A cylindrical neural approximation theorem for conditional laws of McKean-Vlasov equations with common noise ​

Author: Nacira Agram, Reda Hmioui, Jan Rems
Published: 8/11/2026, 4:00:00 AM
Categories: math.PR, cs.LG

arXiv:2608.08040v1 Announce Type: cross Abstract: We introduce conditional cylindrical neural networks for approximating functionals of conditional laws in McKean-Vlasov equations with common noise. Fourier moments of the initial law and truncated signatures of the time augmented common noise are ma...

📖 Read original article


219. NeuroGuard: Neural Gradient Update Aware of Representation Damage ​

Author: Taigo Sakai, Kazuhito Hotta
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.08068v1 Announce Type: cross Abstract: Long-tailed class-incremental learning (LT-CIL) must learn new classes from imbalanced streams while retaining old classes. Existing methods mainly change replay, classifiers, or losses. We study a different factor, namely how strongly the feature re...

📖 Read original article


220. PATH: Next-Interval Prediction via Autoregressive Tree Hierarchy on Tabular Data ​

Author: Pengxiang Cai, Wanchen Lian, Chenyang Liu, Xiaohan Li, Qingyuan Zeng, Jinhong Wang, Jintai Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.08078v1 Announce Type: cross Abstract: Interval prediction aims to achieve a target coverage level while producing intervals that are as short as possible. Many conformal regression pipelines first predict an uncertainty surrogate and then convert it into an interval through calibration o...

📖 Read original article


221. RotaryQuant: Fitting 120B MoE Models on Consumer Hardware via Fused Compressed-Space Attention ​

Author: Anthony. Lui, Mohamed. Elsaied, N. P. Savani
Published: 8/11/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2608.08081v1 Announce Type: cross Abstract: Large mixture-of-experts (MoE) language models with 26--120 billion parameters exceed the memory capacity of consumer devices through three simultaneous pressures: resident weight matrices, key-value (KV) cache state that grows linearly with context,...

📖 Read original article


222. Defending Retrieval-Augmented Intrusion Detection Against Knowledge Poisoning and Prompt Injection ​

Author: Kaysarul Anas Apurba, Md. Hasibul Hasan, Mahedee Zaman Moon, Sk. Md. Mizanur Rahman, Atsuo Inomata
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.08100v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enables large language models to classify network flows and generate human-readable incident reports by retrieving semantically similar historical traffic from a vector knowledge base. However, the retrieval layer...

📖 Read original article


223. Hierarchical Multi-Task Federated Learning in VANETs ​

Author: M. Saeid HaghighiFard, Sinem Coleri
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SY, cs.AI, cs.DC, cs.LG, cs.NI, cs.SY

arXiv:2608.08111v1 Announce Type: cross Abstract: Vehicular Ad hoc Networks (VANETs) increasingly rely on federated learning (FL) to enable collaborative intelligence without sharing raw sensory data. However, most existing vehicular FL frameworks assume that all vehicles train a single global model...

📖 Read original article


224. Finite basis physics-informed neural networks with hard constraints for viscous fluid flow in highly perforated domains ​

Author: Jeeeun Lee, Denis Korolev, Miro Duhovic, Seong Su Kim
Published: 8/11/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG

arXiv:2608.08114v1 Announce Type: cross Abstract: In this work, viscous fluid flow governed by the Stokes equations in highly perforated domains is studied using physics-informed neural networks (PINNs). Perforated microstructures induce complex boundary conditions and fine-scale flow features that ...

📖 Read original article


225. Neurosymbolic Discovery of Algebraic Graph Constructions ​

Author: David Seka, Stefan Szeider
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SC, math.CO

arXiv:2608.08118v1 Announce Type: cross Abstract: There are several methods for searching for graphs with prescribed properties, such as SAT solvers and specialized generators. These methods return the result as raw data: an adjacency matrix or a string encoding. The raw data certifies that the grap...

📖 Read original article


226. EFFEKT: Efficient Federated Knowledge Transfer to Foundation Models ​

Author: Matteo Caligiuri, Francesco Barbato, Pietro Zanuttigh, Francesco Restuccia
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.08138v1 Announce Type: cross Abstract: Recent data protection laws have accelerated the adoption of Federated Learning (FL) for privacy-preserving decentralized training. Nevertheless, increasing model sizes impose substantial computational demands on client devices, limiting FL applicabi...

📖 Read original article


227. A Unified Framework for Dynamic Reward Shaping in Reinforcement Learning ​

Author: Fouad Bahrpeyma
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.08158v1 Announce Type: cross Abstract: Sparse, delayed, and weakly informative rewards remain central obstacles to efficient reinforcement learning. Reward shaping addresses these limitations by supplementing the task reward with an auxiliary signal that can accelerate learning while, in ...

📖 Read original article


228. Wiener Representation Filtering for VLM Hallucination Suppression ​

Author: Ameen Ali, Tamim Zoabi, Lidor Brami, Lior Wolf
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.08167v1 Announce Type: cross Abstract: Vision-language models (VLMs) excel at open-ended captioning and visual QA but often describe objects, attributes, or relations absent from the image, a phenomenon known as object hallucination. We propose a {training-free, post-hoc representation ed...

📖 Read original article


229. Matching Supervision to the Student's Learning Capacity: A Unified Framework for On-Policy Self-Distillation ​

Author: Yongkang Yang, Zhezheng Hao, Hong Zhang, Yi Liu, Xiankun Lin, Wence Ji, Fanjunduo Wei, Jiarui Yu, Qiang Lin, Xiaoyun Liang, Hande Dong
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.08176v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) improves the reasoning abilities of LLMs by internalizing privileged context into model parameters through self-distillation. Two recent research lines promote vanilla OPSD by choosing which tokens to learn from and...

📖 Read original article


230. Quantization Degradation in Large Language Models: A Signal-Noise Perspective ​

Author: Chenxi Zhou, Pengfei Cao, Jinyu Ye, Bohan Yu, Haida Yu, Jiang Li, Jun Zhao, Kang Liu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.08188v1 Announce Type: cross Abstract: Post-training quantization reduces the deployment cost of large language models, yet how severely a quantized model degrades is not determined by bit-width alone. We systematically study weight-only post-training quantization across bit-widths, quant...

📖 Read original article


231. Conditional Diffusion for Nonparametric Instrumental Variable Quantile Regression ​

Author: Xingdong Feng, Xinhong Jiang, Yuling Jiao, Lican Kang, Junwei Liu
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.08204v1 Announce Type: cross Abstract: This work proposes deep nonparametric Instrumental variable quantile regression (IVQR), a two-stage estimator that combines conditional diffusion modeling with a kernel-smoothed conditional moment formulation. In the first stage, we estimate the join...

📖 Read original article


232. On the Robustness of LLMs' Internal Representation of Code Correctness ​

Author: Francisco Ribeiro, Sohaila Abdulsattar, Renata Gonzalez, Mahmoud Kassem, Sarah Nadi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2608.08266v1 Announce Type: cross Abstract: Code generated by modern language models often reads naturally. Yet, it also often fails to implement what was asked. This should be no surprise, as research shows the models' own confidence signals are poorly calibrated with actual correctness. A pr...

📖 Read original article


233. Learning under Opponent Unawareness in Linear-Quadratic Stochastic Games ​

Author: Dantong Chu, Xuefeng Gao, Yufei Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: math.OC, cs.GT, cs.LG, econ.TH

arXiv:2608.08268v1 Announce Type: cross Abstract: As firms increasingly deploy machine learning for strategic decision-making, understanding algorithmic interactions has become central to operations research and economics. This paper studies learning in infinite-horizon, nonzero-sum linear-quadratic...

📖 Read original article


234. Three Necessary Principles for Self-Supervised Visual Representation Learning ​

Author: Nikos Giakoumoglou, Paschalis Giakoumoglou, Tania Stathaki
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.08309v1 Announce Type: cross Abstract: We argue that learning visual representations without labels requires a training signal jointly complete across three non-overlapping objectives: semantic invariance across augmented views, patch-level spatial prediction, and representational non-deg...

📖 Read original article


235. A continually expandable foundation model for brain MRI ​

Author: Michail Mamalakis, Carmen Jimenez-Mesa, Yonghao Li, Hao Chen, Chao Li, Antonios Mamalakis, John Suckling, Richard Bethlehem, Stephen J. Price, Richard J. Gilbertson, Pietro Lio
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.08319v1 Announce Type: cross Abstract: Brain magnetic resonance imaging (MRI) is central to neuroscience and clinical assessment, but models are commonly developed for individual diseases, populations or imaging protocols. Foundation models promise more general representations, yet they a...

📖 Read original article


236. Eikonal Regularisation in Physics-Informed Neural Networks for Three-Dimensional Level-Set Advection: Transferability of Two-Dimensional Design Principles ​

Author: Muhammad Akbar Khan
Published: 8/11/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG, physics.comp-ph

arXiv:2608.08322v1 Announce Type: cross Abstract: Physics-informed neural networks applied to the level-set formulation of interface advection commonly augment the residual and initial-condition losses with an eikonal regulariser, penalising the deviation of $|\nabla\phi|$ from unity. A previous t...

📖 Read original article


237. Physics-Informed Condition Monitoring of SiC Power Modules ​

Author: Mattia Scarpa, Evgeny Kusmenko, Francesco Toso, Mattia Bruschetta, Ruggero Carli, Simon Achatz
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2608.08363v2 Announce Type: cross Abstract: Silicon carbide (SiC) power modules are increasingly deployed in automotive traction inverters, where condition monitoring is essential to prevent in-service failures. Despite extensive qualification under AQG 324, no consolidated approach exists for...

📖 Read original article


238. Failure-Mechanism Transferability of Cumulative-Damage Features for Health State Estimation of SiC Power Modules ​

Author: Mattia Scarpa, Evgeny Kusmenko, Francesco Toso, Mattia Bruschetta, Ruggero Carli, Simon Achatz
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2608.08365v2 Announce Type: cross Abstract: Data-driven health-state estimators for SiC (Silica-Carbide) power modules typically report their performance on a single accelerated-aging campaign, and how that performance transfers to a different failure mechanism is rarely tested. We benchmark f...

📖 Read original article


239. Safety Cost of Steering Vectors Is Separable and Reducible ​

Author: Yuxiao Li, Gjergji Kasneci
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.08383v1 Announce Type: cross Abstract: Steering vectors are a lightweight tool for controlling LLM behavior. However, emerging evidence shows that steering vectors can unintentionally compromise a model's safety mechanisms and increase compliance with harmful requests, while no effective ...

📖 Read original article


240. Does a Toehold Make a Bidder Bolder? Preemption and Multiplicity in Multi-Round Takeover Auctions ​

Author: Zain Naboulsi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.GT, cs.AI, cs.LG

arXiv:2608.08407v1 Announce Type: cross Abstract: A bidder can quietly buy a stake in a company before making an offer for it. That stake, a toehold, is supposed to pay for itself twice: it makes the bidder willing to bid harder, and it frightens rivals into staying out of the fight. The first effec...

📖 Read original article


241. Population-Level Generative Modeling for Ranking Data ​

Author: Zhaoyang Shi
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, math.ST, stat.ML, stat.TH

arXiv:2608.08422v2 Announce Type: cross Abstract: Ranking data arise in scientific and machine learning applications, including recommendation systems, information retrieval, voting, marketing, and AI preference ranking from human feedback. Existing statistical work has primarily focused on inferenc...

📖 Read original article


242. ARC: Augmented-Rank Conformalization for Changepoint Localization --- Finite-Sample Validity and Distribution-Robust Efficiency ​

Author: Chenchen Peng, Mixia Wu, Qijing Yan, Zhiqi Shen, Jie Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.CV, cs.LG

arXiv:2608.08424v1 Announce Type: cross Abstract: Conformal changepoint localization turns any score into a confidence set for the changepoint with finite-sample coverage. Coverage is universal; efficiency is not. The oracle score is a likelihood ratio, so practical scores estimate density ratios, a...

📖 Read original article


243. SuperNeuroMAT: An Efficient Matrix-based Simulator for Spiking Neural Networks ​

Author: Prasanna Date, Kevin Zhu, Shruti Kulkarni, Ashish Gautam, Chathika Gunaratne, Robert Patton, Tyler Nitzsche, Ian Mulet, Zachary Johnson-Scott, Addison Helms, Duncan Rowden, Simon Weston, Maryam Parsa, Catherine Schuman, Thomas Potok
Published: 8/11/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.CE, cs.ET, cs.LG

arXiv:2608.08479v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) offer a promising pathway to energy-efficient AI and brain-inspired computing. However, their widespread adoption is hindered by a lack of fast, accessible, and versatile simulation frameworks. In this paper, we introdu...

📖 Read original article


244. HoloAegis: Frozen Representation, Topological Inference: Minimally Parametric Safety Manifolds for Zero-Shot LLM Guardrails ​

Author: Tak Ho Alex Li, Kaijie Liu, Lik-Hang Lee, Kin Chung Ho, Ping Shum, Michael K. Ng
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.08485v1 Announce Type: cross Abstract: Current LLM safety guardrails face a fundamental tension: fine-tuning distorts pre-trained representations while generative judges incur prohibitive inference costs. We challenge the prevailing paradigm by asking: can safety be achieved through pure ...

📖 Read original article


245. Curriculum Generation under Structured Parametric Environments for Robust Navigation Policies ​

Author: Prishita Ray
Published: 8/11/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.08545v1 Announce Type: cross Abstract: Robust navigation policies for autonomous agents must generalize across continuously varying environmental conditions such as turn rates, obstacles, friction, pits, and slopes. Curriculum generation provides a principled mechanism for improving gener...

📖 Read original article


246. MotionCraft: Latent World Modeling with Sparse Attention for Visual Upscaling ​

Author: Rong Fu, Chunlei Meng, Yangchen Zeng, Xiaowen Ma, Yongtai Liu, Wangyu Wu, Shuo Yin, Zijian Zhang, Sicheng Li, Yingrui Ji, Chenhao Wang, Simon Fong
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.MM

arXiv:2608.08553v1 Announce Type: cross Abstract: Video super-resolution (VSR) aims to recover high-fidelity high-resolution videos from low-resolution inputs and is central to applications ranging from mobile capture to streaming and archival restoration. Existing approaches trade off among local-d...

📖 Read original article


247. Transfer Learning-Enabled Distortion Compensation for Amplitude-Phase-Time Block Modulation-Based Nonlinear Single-Carrier Wireless Communications ​

Author: Guoxing Duan, Min Fan, Cheng Yi, Bensheng Yang, Wei Xu, Haiming Wang, Xiaohu You
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.08554v1 Announce Type: cross Abstract: Power amplifier (PA) nonlinearity and memory effects significantly limit the spectral compliance, reliability, and energy efficiency of communication systems. To address this, we propose a transfer-learning-enabled, fully digital transceiver-cooperat...

📖 Read original article


248. OpenVisTool: An Open Recipe for Synthesizing Instructive Visual Tool-Use Trajectories ​

Author: Changhao Xiang, Shilin Zhang, Zheng Ma, Kanzhi Cheng, Ruize Ma, Yi Feng, Jianbing Zhang, Zhi Wang, Zhen Wu, Xinyu Dai, Lewei Lu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.CV, cs.LG

arXiv:2608.08557v2 Announce Type: cross Abstract: Visual tool use has emerged as a fundamental capability for multimodal agents to actively acquire evidence beyond a fixed image encoding. The prevailing recipe learns this capability from teacher-generated trajectories filtered for answer correctness...

📖 Read original article


249. Differentiate the Solver, Not the Equation: Reverse-Sweep Adjoints for Block Implicit Simulation ​

Author: Lei Shu, Ying Jiang, Kui Wu, Yin Yang, Leonidas Guibas, Chenfanfu Jiang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.GR, cs.LG, cs.NA, math.NA

arXiv:2608.08559v1 Announce Type: cross Abstract: Differentiable simulation is a key component in learning, control, and inverse problems, where gradients through nonlinear implicit solvers are required. Existing approaches either rely on unrolled automatic differentiation, whose memory grows with s...

📖 Read original article


250. LazyHMC: Hamiltonian Monte Carlo Simulation for Lazy, Infinite Dimensional Probabilistic Programs ​

Author: Maria-Nicoleta Cr\u{a}ciun, C. -H. Luke Ong, Tom Schrijvers, Sam Staton
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.PL, stat.CO

arXiv:2608.08588v1 Announce Type: cross Abstract: Hamiltonian Monte Carlo (HMC) is a successful generic inference method in probabilistic programming, but in its ordinary formulation it needs gradients and finite-dimensional parameter spaces. In Haskell, lazy evaluation lets probabilistic programs e...

📖 Read original article


251. Population-Scalable Multi-Agent World Modeling ​

Author: Renjie Zhao, Yuxiang Wu, Mingyu Zhang, Jiaxin Li, Sisi Li, Yimin Sheng, Tianxi Tan, Zhenkai Zhang, Jianyi Zhu, Yong-Lu Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.08600v1 Announce Type: cross Abstract: World models have recently achieved impressive progress in visual prediction and interactive generation, but extending them to multi-agent environments introduces a fundamental scalability challenge. Existing methods generally assume a fixed number o...

📖 Read original article


252. ADEx-FNO: A Unified Ambient-Domain Framework for Fourier Neural Operators on Varying Geometries ​

Author: Roberto Nuca, Giovanni Testa, Luca Galimberti, Matteo Parsani
Published: 8/11/2026, 4:00:00 AM
Categories: math.NA, cs.CE, cs.LG, cs.NA

arXiv:2608.08608v1 Announce Type: cross Abstract: Fourier neural operators (FNOs) provide efficient nonlocal spectral learning, but varying geometries and independently chosen discretizations remain difficult to accommodate. We introduce the ambient-domain extension Fourier neural operator (ADEx-FNO...

📖 Read original article


253. Kernel Methods for Refined Prophet Inequalities ​

Author: Patrick Loiseau, Mathieu Molina, Vianney Perchet, Sebastian Perez-Salazar, Victor Verdugo
Published: 8/11/2026, 4:00:00 AM
Categories: cs.GT, cs.DS, cs.LG, math.OC

arXiv:2608.08662v1 Announce Type: cross Abstract: The single-selection prophet inequality is a canonical Bayesian online selection problem in which independent nonnegative values arrive sequentially and the decision-maker must irrevocably select at most one. Classical single-threshold guarantees are...

📖 Read original article


254. Multi-kernel spectral clustering: Entrywise eigenvector perturbation bounds and exact recovery ​

Author: Zeqin Lin, Guangming Pan, Zhixiang Zhang, Yinbing Zhou
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2608.08704v1 Announce Type: cross Abstract: Kernel spectral clustering with a single bandwidth can be inadequate for data exhibiting multiple characteristic pairwise-distance scales, a problem particularly prevalent in the high-dimensional regime. We address this issue through a multi-kernel f...

📖 Read original article


255. A Distribution Mapping Approach to Counterfactually Fair Reinforcement Learning ​

Author: Jianhan Zhang, Jitao Wang, John D. Piette, Donglin Zeng, Chengchun Shi, Zhenke Wu
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.CY, cs.LG, stat.ME

arXiv:2608.08743v1 Announce Type: cross Abstract: Reinforcement learning (RL) seeks to optimize sequential decisions to maximize population-level benefits over time. However, when deployed in high-stakes settings such as healthcare, RL decisions might systematically restrict some subpopulation's acc...

📖 Read original article


256. Can We Optimize the Performance-Carbon Emission Break-Even Point?: The Quest for Greener LLMs ​

Author: Sourav Das, Tanmay Joshi, Kripabandhu Ghosh
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CE, cs.ET, cs.LG

arXiv:2608.08744v1 Announce Type: cross Abstract: The carbon footprint of any deployed Large Language Model (LLM) accumulates during inference, where repeated use of the model substantially exceeds the one-time cost of fine-tuning. Yet most efficiency interventions target either pre-training scale o...

📖 Read original article


257. A Mean-Field Framework for Inference-Time Distributional Control of Diffusion Models ​

Author: Samuel Howard, Nikolas N"usken
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.08770v1 Announce Type: cross Abstract: Diffusion models are increasingly used as controllable samplers, whose generations can be steered at inference time according to a chosen reward function. While such rewards are typically defined on individual samples, for many applications it is des...

📖 Read original article


258. End-to-End Neural Decomposition with Koopman Operators for Time-Series Forecasting ​

Author: De-Yan Lu, Xugang Lu, Yu Tsao, Jian-Jiun Ding
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.08788v1 Announce Type: cross Abstract: Koopman theory offers a linear-operator view of nonlinear sequence dynamics by lifting observations into a space where evolution is governed by a linear time-invariant Koopman operator. While the Koopman operator provides a linear representation of n...

📖 Read original article


259. ML-Based Hierarchical Prediction for Practical Energy Scheduling in Dynamic NTN-WPT Systems ​

Author: Zhanyu Ju, Wenchi Cheng
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2608.08804v1 Announce Type: cross Abstract: With advancements in long-distance wireless power transfer (WPT) and space-based energy technologies, integrating WPT into non-terrestrial networks (NTNs), referred to as NTN-WPT, is emerging as a promising approach for next-generation wireless netwo...

📖 Read original article


260. 360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents ​

Author: Kenta Watanabe, Atsuyuki Miyai, Mizuki Takenawa, Kiyoharu Aizawa, Toshihiko Yamasaki
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.08814v1 Announce Type: cross Abstract: We present 360CityArena, a benchmark for evaluating the urban exploration capabilities of embodied agents within a photorealistic environment constructed from 360-degree videos. Existing outdoor benchmarks either lack sufficient photorealism or compl...

📖 Read original article


261. Sparse Attention to Emotion: Efficient Facial Emotion Recognition via Token Reduction ​

Author: Aya Manel Zitouni, Aicha Zenakhri, Karim Haroun, Larbi Boubchir
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.08873v1 Announce Type: cross Abstract: Facial Emotion Recognition (FER) is an important task that has significant implications across various fields such as biometrics, health, and human-computer interaction. Current Vision Transformer-based approaches display quadratic complexity $\mathc...

📖 Read original article


262. Inductive Graph Layout with Implicit Neural Fields ​

Author: Berfin Inal, Daniel Probst
Published: 8/11/2026, 4:00:00 AM
Categories: cs.HC, cs.LG

arXiv:2608.08876v1 Announce Type: cross Abstract: A graph layout is normally a table of $N$ free coordinates. We optimise a function with a fixed number of parameters instead. This gives a drawing a sample complexity and an extensible domain. Force-directed algorithms remain the standard tools for g...

📖 Read original article


263. From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability ​

Author: Alexander Hackett, Arnaud Denis-Remillard, Axel Cassou
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.08904v1 Announce Type: cross Abstract: How much of a vision-language model's (VLM) spatial understanding remains after the action post-training process of building a vision-language-action model (VLA)? We probe depth perception, a primitive of spatiogeometric understanding, from every dec...

📖 Read original article


264. Tied Trit-Planes: Constraining PTQTP to a Uniform Nine-Level Quantizer, with a Persistent Folded Format for Disk-Streamed Mixture-of-Experts Serving ​

Author: Matteo Grella
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.08910v1 Announce Type: cross Abstract: PTQTP decomposes LLM weight matrices into two ternary (trit) planes with two free per-group scales. Tying the scales to a fixed ratio of three collapses the decomposition into a single uniform nine-level quantizer, a known balanced-ternary identity. ...

📖 Read original article


265. Physics-Informed Learning for Robust Acoustic Localization with Calibrated Uncertainty ​

Author: Jennifer N. Kampe, Changwoo J. Lee, Xin Shen, Ari Lehti"o, Sandro von Brandenburg, Ossi Nokelainen, David B. Dunson, Otso Ovaskainen
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.SD

arXiv:2608.08911v1 Announce Type: cross Abstract: Recent advances in Passive Acoustic Monitoring (PAM) offer an opportunity to obtain ecological spatial point-process data at unprecedented scale. However, realizing this opportunity necessitates the development of accurate and scalable localization m...

📖 Read original article


266. Clustered Attractor Manifolds and Dynamical Condensation in Self-Attention ​

Author: Qucheng Gao, Zuyi Yang, Xiao Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cond-mat.stat-mech, cs.LG

arXiv:2608.08922v1 Announce Type: cross Abstract: Transformer layers generate state-dependent interaction networks: token representations determine the attention matrix, which in turn updates the representations. We study this feedback in a minimal normalized self-attention dynamics and identify the...

📖 Read original article


267. Decoding Phenotypes: A Framework for Fusing Genomic Language Models and Neuroimaging ​

Author: Tianli Tao, Ziyang Wang, Emma Robinson, Rachel Sparks, Le Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.08926v1 Announce Type: cross Abstract: Neuroimaging and genetic testing are two important clinical references for nervous system diseases, offering complementary diagnostic information. However, integrating genomic and neuroimaging data for precise disease diagnosis is challenging due to ...

📖 Read original article


268. What Would Fix This RAG Failure? Auditing Counterfactual Response with Paired Evidence Interventions ​

Author: Wenzhang Du
Published: 8/11/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.08944v1 Announce Type: cross Abstract: A failed retrieval-augmented generation (RAG) answer can be consistent with several unseen responses to evidence repair. We introduce Pair-ID, an offline audit that holds one query, retrieval state, and reader constant, then crosses two operations, a...

📖 Read original article


269. Can Webcam Gaze Constrain Mesa-Objectives in Driving Models? An Instrument Precision Analysis ​

Author: Lennox Anderson, Ahmed Boutar, Jonah Mulcrone, Tal Erez
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.HC, cs.LG

arXiv:2608.08947v1 Announce Type: cross Abstract: Current hazard detection systems in autonomous driving may develop mesa objectives, learned internal goals that achieve high training performance through spurious correlations rather than genuine hazard recognition. We investigate whether human gaze ...

📖 Read original article


270. Do AI Forecast Ensembles Sample the Correct Conditional Distribution? ​

Author: Lucas J. Howard, Elizabeth A. Barnes
Published: 8/11/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.AI, cs.LG, stat.AP, stat.ML

arXiv:2608.08954v1 Announce Type: cross Abstract: Ensemble forecasting aims to sample the conditional distribution of outcomes; whether AI forecast ensembles do this correctly in a joint sense remains largely untested. We train a diffusion model for probabilistic subseasonal coastal sea level foreca...

📖 Read original article


271. Fourier Self-Supervision for Fine-Grained Generalized Category Discovery ​

Author: Sarah Rastegar, Mina Ghadimi Atigh, Pascal Mettes, Yuki M. Asano, Cees G. M. Snoek
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.08963v1 Announce Type: cross Abstract: Generalized Category Discovery aims to recognize known categories while identifying novel ones within unlabeled data. Existing methods, typically based on self-supervision and contrastive learning, often struggle to capture fine-grained distinctions,...

📖 Read original article


272. Guardian Crawler: Retrieval-First Knowledge Discovery with Bounded LLM Augmentation for Noisy Web Intelligence ​

Author: Joshua Castillo, Santosh Nukavarapu, Ravi Mukkamala
Published: 8/11/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG

arXiv:2608.08994v1 Announce Type: cross Abstract: Retrieving relevant evidence from noisy web data is challenging, particularly in sensitive domains containing incomplete reports, heterogeneous language, and irrelevant content. We present Guardian Crawler, a reproducible retrieval-first testbed for ...

📖 Read original article


273. Mind the Hook: Source-Level Auditing of Privacy Defenses in Retrieval-Augmented Generation ​

Author: Yanhang Li, Zhichao Fan, Zexin Zhuang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.09001v1 Announce Type: cross Abstract: Black-box privacy scores for retrieval-augmented generation (RAG) are difficult to interpret unless the audited defense's active pipeline hook is known. We propose an active-path audit: inventory source-level hooks over retrieval, retrieved content, ...

📖 Read original article


274. A Tight Lower Bound for Smooth Nonconvex Stochastic Optimization with Bounded Gradient Noise ​

Author: Jikai Jin
Published: 8/11/2026, 4:00:00 AM
Categories: math.OC, cs.CC, cs.LG

arXiv:2608.09004v1 Announce Type: cross Abstract: We prove a sharp lower bound for smooth nonconvex stochastic optimization with uniformly bounded gradient noise. In the (K=1) fresh-sample model, every randomized adaptive algorithm requires $$\Omega\left( \frac{\Delta L}{\epsilon^2} + \frac{\Delta...

📖 Read original article


275. PreGress: Ranking-Native Pre-training and Prompting for Graph Node Ranking ​

Author: Lujie Ban, Jiasheng shi, Yingli Zhou, Kaiwen Xue, Daiyin Wang, Xubin Li, Shuanghua Li, Chenhao Ma
Published: 8/11/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.09016v1 Announce Type: cross Abstract: Node ranking is a fundamental problem in graph information retrieval, measuring the relative importance of nodes and supporting a wide range of applications such as influence analysis, recommendation, and graph-based retrieval augmented generation. H...

📖 Read original article


276. Closing the loop in learning with missing data ​

Author: Dimitrios Pylorof, Humberto E. Garcia
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC

arXiv:2608.09030v1 Announce Type: cross Abstract: What should a machine learning model learn when data is missing during training? We look at the learning process from a dynamical systems perspective, cast data missingness as a structured loss of actuation that limits controllability of the paramete...

📖 Read original article


277. Decision-Focused Learning in Network Interdiction Games ​

Author: Luca M. Hartmann, Parinaz Naghizadeh
Published: 8/11/2026, 4:00:00 AM
Categories: cs.GT, cs.LG, math.OC

arXiv:2608.09036v1 Announce Type: cross Abstract: We study decision-focused learning (DFL) in shortest-path network interdiction (SPNI) games, a Stackelberg game where an interdictor (leader) strengthens the networks' arcs against attacks, while an evader (follower) who is uncertain about costs of a...

📖 Read original article


278. Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations ​

Author: Hongxiang Gao, He-yang Xu, Yuwen Li, Minghui Zhao, Zhipeng Cai, Xingyao Wang, Chenxi Yang, Jianqing Li, Chengyu Liu
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SP, cs.CV, cs.LG

arXiv:2608.09053v1 Announce Type: cross Abstract: Cardiologists interpret electrocardiograms by localizing waveform components, measuring rhythm and interval patterns, and translating these structured observations into diagnostic evidence. Whether this expert reading process can serve as an effectiv...

📖 Read original article


279. Personalized Federated Learning via Variance-Aware Nonparametric Empirical Bayes ​

Author: Jae Ho Chang, Arnab Auddy, Subhadeep Paul
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.09074v1 Announce Type: cross Abstract: We develop a new approach to Personalized Federated Learning across heterogeneous clients using Nonparametric Empirical Bayes (NPEB). Leveraging the asymptotic normality of local parameter estimates obtained from Empirical Risk Minimization or M-esti...

📖 Read original article


280. When Confidence Fails: Overconfidence in LLMs under Uncertainty and Missing Clinical Information ​

Author: Maryam Tahermazandarani, Adnan Mahmood, Fahmida Islam, Quan Z. Sheng
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC, cs.LG

arXiv:2608.09080v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved strong performance in medical question answering and clinical reasoning tasks. However, their reliability under uncertainty remains poorly understood which raises critical concerns for deployment in high-sta...

📖 Read original article


281. A Multi-Scale Temporal Framework with Dynamic Fusion for EEG-Based Emotion Recognition ​

Author: Stefanos Gkikas, Yang Guo, Guangliang Li, Raul Fernandez Rojas, Giorgos Giannakakis, Randy Gomez
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SD

arXiv:2608.09088v1 Announce Type: cross Abstract: Mixed emotions represent a clinically relevant but still underexplored target for automatic emotion recognition. EEG provides millisecond-level access to neural activity, yet most EEG pipelines analyze the signal through a single temporal window, the...

📖 Read original article


282. Contrastive Mask Fidelity: Reference-Free Auditing of Ground-Truth Masks in Remote Sensing Semantic Segmentation ​

Author: Shuaishuai Cao, Shuwei Peng, Meng Tang, Min Huang, Youjin Wang, Jie Chen, Jing Ouyang, Zhiwei Zhai
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.09101v1 Announce Type: cross Abstract: Semantic segmentation models are trained and evaluated against human-drawn masks, yet remote-sensing annotations are often coarse, incomplete, or misaligned; high overlap scores may then reflect agreement with imperfect labels rather than faithfulnes...

📖 Read original article


283. Multitask Scanning Probe Microscopy ​

Author: Aditya Raghavan, Yu Liu, Ian Mercer, JP Maria, Sergei Kalinin
Published: 8/11/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG, physics.ins-det

arXiv:2608.09104v1 Announce Type: cross Abstract: Scanning probe microscopy provides nanoscale access to structural, electrical, electromechanical, magnetic, and mechanical properties of materials. Its increasing use for wafer-scale characterization and combinatorial materials exploration creates a ...

📖 Read original article


284. TRACE: TRajectory Attribution for Automated Context Engineering ​

Author: Yikai Zhao, Pradeep Kumar Misra, Saurabh Pandey
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.09153v1 Announce Type: cross Abstract: Production AI agents fail when their context sources -- system prompts, knowledge bases, tool descriptions, and procedural skills -- contain errors or gaps. Current maintenance relies on manual log review and ad-hoc debugging, creating a scalability ...

📖 Read original article


285. Particle-Based Conformal Prediction for Contact-Aware Uncertainty Calibration in Stratified Configuration Spaces ​

Author: Lu'is Marques, Kristian Popov, Dmitry Berenson
Published: 8/11/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, stat.ME

arXiv:2608.09166v1 Announce Type: cross Abstract: Reliable uncertainty representation is essential for deploying autonomous systems that interact with their environment, as robots must reason about how uncertainty arising from both stochasticity and model mismatch is impacted by contacts with obstac...

📖 Read original article


286. Rethinking Medical Landmark Localization with Prototype Learning-based Progressive Offset Correction ​

Author: Jingxian Xu, Yuhao Huang, Rusi Chen, Yanfeng Zhou, Dong Ni
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.09182v1 Announce Type: cross Abstract: Accurate landmark localization in medical images is a fundamental step for quantitative clinical measurement and downstream analysis. Existing localization methods have advanced, among which multi-stage refinement is a superior solution. Although thi...

📖 Read original article


287. CPDA: Class-Conditional Path Distribution Alignment for Unsupervised Time-Series Domain Adaptation ​

Author: Felix Ott, Christopher Mutschler
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.09193v1 Announce Type: cross Abstract: Unsupervised time-series domain adaptation (DA) addresses the challenge of transferring a classifier from a labeled source domain to an unlabeled target domain under distribution shifts induced by different users, sensors, devices, acquisition condit...

📖 Read original article


288. UNMASK: Discovering and Causally Verifying Spurious Shortcuts in Text Classifiers ​

Author: Chidaksh Ravuru, Shashank Srivastava
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.09209v1 Announce Type: cross Abstract: Neural language models trained on large crowdsourced corpora frequently exploit spurious surface patterns tied to target labels without true linguistic or causal relevance, boosting benchmark performance while failing on adversarial or out-of-distrib...

📖 Read original article


289. MoRSE: Task-Oriented Multi-Agent System with Mixture of Role-Subtask Experts ​

Author: Peiwen Li, Shiyang Zhang, Yangtian Zhang, Sizhuang He, David van Dijk, Rex Ying
Published: 8/11/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.CL, cs.LG

arXiv:2608.09251v1 Announce Type: cross Abstract: Large language model-based multi-agent systems have recently shown strong potential for complex, long-horizon tasks. However, existing methods mainly rely on coarse prompt-level differentiation without parameter adaptation for diverse subtasks, resul...

📖 Read original article


290. Privileged Likelihood Is Not Automatically Value: Three Checks for Token Credit in On-Policy Self-Distillation ​

Author: Xuan-Phi Nguyen, Shrey Pandit, Yiran Zhao, Anurag Koul, Zeyu Liu, Shafiq Joty
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.09263v1 Announce Type: cross Abstract: Outcome verifiers score completed reasoning traces but do not assign credit to intermediate tokens. Privileged self-distillation attempts to fill this gap by rescoring a model's own rollout with training-only information. A token likelihood change, h...

📖 Read original article


291. Did the Grid Erase the Event? EndoClock for Auditing Medical World-Model Pipelines ​

Author: Yarin Udi, Tom Sharon-Shahak, Roee Masad, Dan Pri-Tal
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.09266v1 Announce Type: cross Abstract: Medical world models commonly learn from multimodal recordings synchronized onto a fixed-rate grid. This preprocessing resamples each native stream onto a shared time axis. Each stream has an observation clock that governs when observations are emitt...

📖 Read original article


292. Verifiably grounded machine interpretation of lunar geology ​

Author: Tom Sander, Kay Wohlfarth, Christian W"ohler
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.09276v1 Announce Type: cross Abstract: Planetary geology relies on historical, interpretive reasoning to reconstruct past events from diverse observations. Here, we present a step toward an automated "machine intelligence geologist" by embedding this distinct methodology of geologic knowl...

📖 Read original article


293. CADEngBench: It Looks Like CAD, but Does It Work? Evaluating Parametric Design, Assembly Reasoning, and Physics Simulation ​

Author: Harmanjot Singh, Abhra Dubey, Jorge Alejandro Amador Herrera
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG, cs.RO

arXiv:2608.09296v1 Announce Type: cross Abstract: A CAD model is not engineering-grade merely because it looks correct. It must satisfy design requirements, respond predictably to parameter changes, support controlled edits, match a reference structural response under a declared analysis, and connec...

📖 Read original article


294. SAFE-CHEM: Uncertainty-Aware Policy Switching for Robust Robotic Chemistry ​

Author: Laura Jones, Shazil Shahzad, Ayesha Sana, Gabriella Pizzuto
Published: 8/11/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.09303v1 Announce Type: cross Abstract: The deployment of autonomous robotic systems in chemistry laboratories is accelerating experimental workflows and providing the foundational data for AI-driven scientific discovery. However, despite the success of data-driven methods in acquiring dex...

📖 Read original article


295. Control-Oriented Scenario Tree Construction through Reinforcement Learning ​

Author: Fabio Pavirani, Bert Claessens, Pierre Pinson, Chris Develder
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SY, eess.SY

arXiv:2608.09335v1 Announce Type: cross Abstract: Multistage stochastic model predictive control (MPC) handles uncertainty by optimizing over a scenario tree, a finite branching approximation of future outcomes constructed from sampled forecasts. To build such a tree, conventional methods focus on m...

📖 Read original article


296. A Machine Learning Based Search for Lunar Anomalies ​

Author: Cameron Kelahan, Daniel Angerhausen, Adam Lesnikowski, Valentin T. Bickel
Published: 8/11/2026, 4:00:00 AM
Categories: astro-ph.EP, cs.LG

arXiv:2608.09350v1 Announce Type: cross Abstract: The Lunar Reconnaissance Orbiter (LRO) has been collecting high-resolution images (at around 0.5-2 meters per pixel linearly with its Narrow Angle Camera) of the Moon since 2009, amassing a large dataset of images and offering researchers the opportu...

📖 Read original article


297. Deep Learning based Detection of Fishing Vessels and Fishing Monitoring using Nightlight Images ​

Author: Shantakar Mohanty, Prasun Kumar Gupta, Raian Vargas Maretto
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV, stat.AP

arXiv:2608.09360v1 Announce Type: cross Abstract: The demand for maritime surveillance has given rise to the need for monitoring fishing vessel activities, particularly in addressing the challenge of "dark vessels" that operate without Automatic Identification System (AIS) transmission. This study p...

📖 Read original article


298. Coordinate-Residual Physics-Driven Neural Network for Electromagnetic Inverse Scattering ​

Author: Yutong Du, Zicheng Liu, Bo Qi, Yali Zong, Peixian Han
Published: 8/11/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG, physics.app-ph

arXiv:2608.09382v1 Announce Type: cross Abstract: Electromagnetic inverse scattering is a nonlinear and ill-posed problem, where accurate reconstruction is challenging due to measurement limitations, noise, and high computational costs, especially for 3-D imaging. Although physics-driven neural netw...

📖 Read original article


299. Regret, equilibrium, and learning in games: A guided tour ​

Author: Panayotis Mertikopoulos
Published: 8/11/2026, 4:00:00 AM
Categories: cs.GT, cs.LG, math.OC

arXiv:2608.09389v1 Announce Type: cross Abstract: This note aims to serve as an entry point to the literature on learning in games, a topic with significant theoretical appeal and a wide range of applications -- from machine learning and data science to economics and beyond. Our presentation is stru...

📖 Read original article


300. Walk-on-Spheres Monte Carlo and deep neural network approximations of elliptic PDEs with drift and killing ​

Author: Konrad Kleinberg, Thomas Kruse
Published: 8/11/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.AP, math.PR

arXiv:2608.09494v1 Announce Type: cross Abstract: In this paper we provide Monte Carlo and deep neural network approximations for stochastic representations of solutions to linear elliptic partial differential equations with constant diffusion, drift and killing. Building on the modified Walk-on-Sph...

📖 Read original article


301. XFeat Revisited: Reproducibility and Evaluation of a Lightweight Image Matcher ​

Author: Lazar {\DJ}okovi'c, Aimee Lin
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.09519v1 Announce Type: cross Abstract: We present a reproducibility study of XFeat, a lightweight local feature extractor and matcher designed to identify corresponding points across images efficiently on resource-constrained hardware. We re-implement the architecture based on the paper a...

📖 Read original article


302. Distributed Optimization with Streaming Data: A Temporal Weighting Perspective ​

Author: Muhammad Faraz Ul Abrar, Nicol`o Michelusi, Erik G. Larsson
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG, cs.SY, eess.SY, math.OC

arXiv:2608.09565v1 Announce Type: cross Abstract: Optimization theory is a widely used tool for intelligent decision-making. While classical optimization deals with fixed, time-invariant objective functions, many modern applications operate in dynamic environments where data arrive sequentially, and...

📖 Read original article


303. Structure-Enhanced Features and Quality-Aware Dynamic Anchor Scoring for Robust Lane Detection ​

Author: Weize Cai, Yongqi Dong, Zhida Shao, Yichen Liu, Zixin Fu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV, eess.SP

arXiv:2608.09610v1 Announce Type: cross Abstract: Lane detection requires recovering thin, elongated, and frequently occluded lane structures under challenging driving conditions. While anchor-based detectors provide efficient candidate generation, their performance is limited by two coupled issues:...

📖 Read original article


304. LoRA-based Adaptation Alone Is Not Enough: Understanding the Limits of Foundation Models for Face Presentation Attack Detection ​

Author: Peter Lorenz, Anjith George, Marcel S'ebastien
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.09633v1 Announce Type: cross Abstract: Face presentation attack detection (PAD) aims to reliably detect a wide range of presentation attacks. While PAD methods achieve strong performance within individual datasets, their performance degrades under cross-dataset evaluation. Variations in s...

📖 Read original article


305. Activation Probes Surface Code-Security Signals that the Model's Output Misses ​

Author: Ivan Wiryadi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.09643v1 Announce Type: cross Abstract: AI coding agents now write a growing share of production code, and human security review does not scale at the rate code is generated. The agents in widest use are closed-weight, so a deploying team cannot read their internals. It can instead run an ...

📖 Read original article


306. Test-Time Scaling for CAD Generation via Verifier-Free Consensus Selection ​

Author: Aaron Haag, Altay Ka\c{c}an, Bertram Fuchs, Oliver Lohse
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CE, cs.LG

arXiv:2608.09706v1 Announce Type: cross Abstract: Large language models can write parametric CAD programs from a natural-language description (text-to-CAD generation), but a single sample is often wrong. Increasing test-time compute by sampling multiple candidates only helps if a good candidate can ...

📖 Read original article


307. Input convex neural networks as surrogates in mathematical optimisation ​

Author: Yu Liu, Jan Kronqvist, Fabricio Oliveira
Published: 8/11/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2608.09707v1 Announce Type: cross Abstract: Embedding trained neural networks as surrogates within optimisation problems is an established practice in operations research. The prevailing approach uses feedforward neural networks (FNNs) with ReLU activations, whose piecewise-linear structure ad...

📖 Read original article


308. Defining Decentralization: An Ontological Perspective ​

Author: Jakub Kacper Szel\k{a}g, Aydin Abadi, Mohammad Naseri
Published: 8/11/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG, cs.LO, cs.SY, eess.SY

arXiv:2608.09748v1 Announce Type: cross Abstract: Decentralization as a concept in computer science has existed for over half a century. Despite its fundamental role across domains such as security, distributed computing, artificial intelligence, cloud infrastructures, and Internet of Things (IoT) a...

📖 Read original article


309. Disentangling Co-Occurring Retinal Pathologies with Saliency-Guided Sparse Expert Routing ​

Author: Nagur Shareef Shaik, Jeongwoo Park, Yeong-Jin Kim, Jaeuk Jung, Hyunjung Oh, Dong Hye Ye
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV, eess.SP

arXiv:2608.09752v1 Announce Type: cross Abstract: Retinal fundus images frequently exhibit multiple co-occurring pathologies, yet standard deep learning classifiers apply static, identical computation to every image regardless of the underlying disease distribution. We propose a novel architecture t...

📖 Read original article


310. C$^2$A: Coupling Spatial Evidence with Clinical Priors via Co-occurrence Aware Class Attention for Multi-Label Chest X-Ray Classification ​

Author: Akash Gogineni, Nagur Shareef Shaik, Aasrith Mandava, Adnan Masood, Dong Hye Ye
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV, eess.SP

arXiv:2608.09774v1 Announce Type: cross Abstract: Thoracic pathologies rarely occur in isolation, yet standard multi-label classifiers rely on shared global descriptors, discarding \emph{where} findings lie and \emph{how} they co-occur. We propose \textbf{C$\mathbf{^2}$A} (Co-occurrence Aware Class ...

📖 Read original article


311. AirFlow: Context Preserving and Multi-Rate State Modeling for Air Quality Forecasting ​

Author: Fan Yang, Nan Chen, Yijie Dong, Yuchen Zhang, Wei Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CE, cs.LG

arXiv:2608.09775v1 Announce Type: cross Abstract: Accurate air quality forecasting is essential for public health and urban environmental management, but remains challenging because pollutant channels differ in periodicity and distribution drift, while their concentration trajectories contain both m...

📖 Read original article


312. RA-FinBERT: Rule-aware LoRA adaptation for low-resource financial sentiment classification ​

Author: Fan Zhang, Jiaming Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.09834v1 Announce Type: cross Abstract: Financial sentiment analysis converts unstructured financial news into quantitative signals that can support market analysis and decision-making. Existing work on resource-efficient financial NLP has largely focused on compressing or adapting pretrai...

📖 Read original article


313. RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance ​

Author: Dongchi Huang, Hongyin Zhang, Bohan Hou, Siteng Huang, Zhian Su, Hang Guo, Tong Lu, Zhaofeng Xu, Jiahao Tang, Jianfei Yang, Donglin Wang, Peixi Peng, Mingxiu Chen, Deli Zhao, Xin Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2608.09853v1 Announce Type: cross Abstract: General-purpose reward models are increasingly the bottleneck for scaling robot learning, yet the recipe for learning value-related capabilities from large-scale heterogeneous corpora remains underexplored. Existing approaches tie supervision to task...

📖 Read original article


314. Stealing Reasoning Traces from Proprietary LLM APIs ​

Author: Alexander Panfilov, David Schmotz, Ilia Shumailov, Luca Beurer-Kellner, Joachim Schaeffer, Ameya Prabhu, Jonas Geiping, Maksym Andriushchenko
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and limit information leakage. Rather than storing these traces server-side, providers return them to the c...

📖 Read original article


315. Logarithmic-Free Moment and Generalization Bounds for Uniformly Stable Algorithms ​

Author: Thanh Nguyen-Cung, Binh T. Nguyen
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR, math.ST, stat.TH

arXiv:2608.09870v1 Announce Type: cross Abstract: Uniform stability is a classical tool for controlling the generalization error of a learning algorithm. Bousquet, Klochkov, and Zhivotovskiy (2020) showed that the problem can be reduced to a moment inequality for a sum of weakly interacting function...

📖 Read original article


316. Financial Numerical Prediction and Allocation as Token Generation ​

Author: Xu Ouyang, Moontae Lee
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.09880v1 Announce Type: cross Abstract: Financial prediction typically relies on task-specific regression, ranking, or policy heads, separating the language model from the numerical object ultimately evaluated. We investigate whether a causal language model can instead represent forecasts ...

📖 Read original article


317. Space-Creating versus Dead Possession: An Off-Ball Possession-Quality Index for Broadcast Football ​

Author: Seongjin Choi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.CY, cs.LG

arXiv:2608.09887v1 Announce Type: cross Abstract: Ball possession is the most-cited and most-misleading number in football: 60% recycled in one's own half is not 60% spent pinning the opponent back. Existing event-based possession-value frameworks (expected threat, VAEP, on-ball value) price on-ball...

📖 Read original article


318. BDH-CQ: In-Context Learning with Recurrent Latent Reasoning ​

Author: Bj"orn Engdahl, Adrian Kosowski, Jan Chorowski, Zuzanna Stamirowska, Przemys{\l}aw Uzna'nski, Junlin Jiang, Rohan Phadke, Remigiusz Kinas, Richard Zhong
Published: 8/11/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG, stat.ML

arXiv:2608.09888v1 Announce Type: cross Abstract: We introduce BDH-CQ, a reasoning model that combines in-context learning with recurrent latent reasoning. Inputs presented at inference time continuously update the model's recurrent memory; the model then solves a query through iterative computation...

📖 Read original article


319. Consilience for Verifier-Free Test-Time Scaling ​

Author: Lecheng Kong, Like Hui, Haitao Mao, Jun Huan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.09898v1 Announce Type: cross Abstract: Test-time scaling often uses an external verifier, such as compilers and test cases in coding or trained value functions in robotics applications, to obtain high-quality rollouts. Verifier-free test-time scaling (or VF-TTS) is gaining extensive atten...

📖 Read original article


320. Multimodal Model Diffing for Feature Discovery and Control ​

Author: Hunar Batra, Lachin Naghashyar, Ashkan Khakzar, Philip Torr, Christian Schroeder de Witt, Constantin Venhoff, Ronald Clark
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG

arXiv:2608.09928v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) exhibit strong visual understanding, yet the internal features that cause these behaviors remain difficult to identify, audit, or control. While applicable to post-hoc inspection, hidden states that are decomp...

📖 Read original article


321. Understanding Alternating Minimization for Matrix Completion ​

Author: Moritz Hardt
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, stat.ML

arXiv:1312.0925v4 Announce Type: replace Abstract: Alternating Minimization is a widely used and empirically successful heuristic for matrix completion and related low-rank optimization problems. Theoretical guarantees for Alternating Minimization have been hard to come by and are still poorly under...

📖 Read original article


322. Weak Correlations as the Underlying Principle for Linearization of Gradient-Based Learning Systems ​

Author: Ori Shem-Ur, Yaron Oz
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, hep-th, math.PR, stat.ML

arXiv:2401.04013v2 Announce Type: replace Abstract: Deep learning models, such as wide neural networks, can be conceptualized as nonlinear dynamical physical systems characterized by a multitude of interacting degrees of freedom. Such systems in the infinite limit, tend to exhibit simplified dynamic...

📖 Read original article


323. Machine Learning and Data Analysis Using Posets: A Survey ​

Author: Arnauld Mesinga Mwafise
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2404.03082v3 Announce Type: replace Abstract: Partially ordered sets (posets) are discrete mathematical structures that formalize the notion of comparison without forcing every pair of objects to be comparable. This makes them a natural representation for the many machine learning and data-ana...

📖 Read original article


324. Transformer Explainer: Learning LLM Transformers with Interactive Visual Explanation and Experimentation ​

Author: Aeree Cho, Grace C. Kim, Alexander Karpekov, Seongmin Lee, Alec Helbling, Benjamin Hoover, Zijie J. Wang, Minsuk Kahng, Duen Horng Chau
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.HC

arXiv:2408.04619v2 Announce Type: replace Abstract: The Transformer architecture underpins modern large language models powering state-of-the-art text generation and AI applications. However, its complexity makes it difficult for non-experts to learn. Existing resources often lack interactivity, rel...

📖 Read original article


325. See Me, Believe Me: Causality, Intersectionality, and Interventions Improving the Appearance of Patients ​

Author: Kenya S. Andrews, Mesrob I. Ohannessian, Elena Zheleva
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2410.01227v2 Announce Type: replace Abstract: In the context of medical records, patients often experience testimonial injustice, where the textual account undermines the validity of their experiences. Past work has demonstrated that intersectionality of demographic features is crucial to \emp...

📖 Read original article


326. Regret of exploratory policy improvement and $q$-learning ​

Author: Wenpin Tang, Xun Yu Zhou
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, math.OC, math.PR

arXiv:2411.01302v2 Announce Type: replace Abstract: We study the convergence of $q$-learning and related algorithms introduced by Jia and Zhou (J. Mach. Learn. Res., 24 (2023), 161) for controlled diffusion processes. For exploratory policy improvement, we establish exponential convergence under gro...

📖 Read original article


327. ProPINN: Demystifying Propagation Failures in Physics-Informed Neural Networks ​

Author: Yuezhou Ma, Haixu Wu, Hang Zhou, Huikun Weng, Jianmin Wang, Mingsheng Long
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2502.00803v3 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) have earned high expectations in solving partial differential equations (PDEs), but their optimization usually faces thorny challenges due to the unique derivative-dependent loss function. By analyzing the l...

📖 Read original article


328. On the Effect of Sampling Diversity in Scaling LLM Inference ​

Author: Tianchun Wang, Zichuan Liu, Yuanzhou Chen, Jonathan Light, Weiyang Liu, Haifeng Chen, Xiang Zhang, Wei Cheng
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2502.11027v5 Announce Type: replace Abstract: Large language model (LLM) scaling inference is key to unlocking greater performance, and leveraging diversity has proven an effective way to enhance it. Motivated by the observed relationship between solution accuracy and meaningful response diver...

📖 Read original article


329. Advancing Intelligent Sequence Modeling: Evolution, Trade-offs, and Applications of State-Space Architectures from S4 to Mamba ​

Author: Shriyank Somvanshi, Md Monzurul Islam, Mahmuda Sultana Mimi, Sazzad Bin Bashar Polock, Gaurab Chhetri, Anandi Dutta, Amir Rafe, Subasish Das
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2503.18970v4 Announce Type: replace Abstract: Structured State Space Models (SSMs) have become a prominent class of sequence models, developed against two long-standing difficulties: the sequential computation and gradient propagation limits of Recurrent Neural Networks (RNNs), and the quadrat...

📖 Read original article


330. Analogical Learning for Cross-Scenario Generalization: Framework and Application to Intelligent Localization ​

Author: Zirui Chen, Hongning Ruan, Zhaoyang Zhang, Ziqing Xing, Ridong Li, Zhaohui Yang, M'erouane Debbah
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, eess.SP

arXiv:2504.08811v3 Announce Type: replace Abstract: Modern learning systems often struggle with joint learning across diverse scenarios and immediate adaptation to new ones, because they rely heavily on the scenario-dependent absolute data-label representations. Here, we propose analogical learning ...

📖 Read original article


331. Decentralized Nonconvex Composite Federated Learning with Gradient Tracking and Momentum ​

Author: Yuan Zhou, Xinli Shi, Xuelong Li, Jiachen Zhong, Guanghui Wen, Jinde Cao
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, math.OC

arXiv:2504.12742v2 Announce Type: replace Abstract: Decentralized Federated Learning (DFL) enables collaborative model training without relying on a central server. When local objectives are nonconvex and coupled with nonsmooth weakly convex regularization, DFL gives rise to a challenging decentrali...

📖 Read original article


332. An Expectation-Maximization Perspective on Reinforcement Learning for LLM Reasoning ​

Author: Tianbing Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2504.18587v2 Announce Type: replace Abstract: Reinforcement learning has emerged as a powerful approach for improving the reasoning capabilities of large language models, as demonstrated by systems such as OpenAI's O1~\cite{o1} and DeepSeek-R1~\cite{r1}. However, widely used algorithms such as...

📖 Read original article


333. FoMoH: A clinically meaningful foundation model evaluation for structured electronic health records ​

Author: Vincent Jeanselme, Zilin Jing, Aparajita Kashyap, Chao Pang, Florent Pollet, Young Sang Choi, Xinzhuo Jiang, Yuta Kobayashi, Yanwei Li, Sara Matijevic, Karthik Natarajan, Shalmali Joshi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2505.16941v4 Announce Type: replace Abstract: Foundation models (FMs) promise to address core limitations of traditional supervised machine learning: (i) reliance on large amounts of labeled data, (ii) task specificity, and (iii) poor transportability. Despite methodological advances in struct...

📖 Read original article


334. The Cell Must Go On: Agar.io for Continual Reinforcement Learning ​

Author: Mohamed A. Mohamed, Kateryna Nekhomiazh, Vedant Vyas, Marcos M. Jose, Andrew Patterson, Marlos C. Machado
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2505.18347v3 Announce Type: replace Abstract: Continual reinforcement learning (RL) concerns agents that are expected to learn continually, rather than converge to a policy that is then fixed for evaluation. This setting is well-suited to environments that the agent perceives as changing over ...

📖 Read original article


335. An Information-Theoretic Framework for Feature Construction in Out-of-Distribution Detection ​

Author: Sudeepta Mondal (Mary), Xinyi (Mary), Xie, Alex Wong, Ganesh Sundaramoorthi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2506.14194v2 Announce Type: replace Abstract: We present a theory for the construction of out-of-distribution (OOD) detection features for neural networks. We introduce random features for OOD through a novel information-theoretic loss functional consisting of two terms, the first based on the...

📖 Read original article


336. Transformer Circuits Can Realize Clustering Algorithms ​

Author: Kenneth L. Clarkson, Lior Horesh, Takuya Ito, Charlotte Park, Parikshit Ram
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2506.19125v2 Announce Type: replace Abstract: Although transformers are most commonly optimized as statistical sequence models, it is unclear to what extent they can implement and learn exact algorithmic computations. Here, we specify a transformer implementation from first principles that exe...

📖 Read original article


337. TS-Mob: Social and Geographical-Aware Time Series Foundation-Model Framework for Human Mobility Prediction ​

Author: Massimiliano Luca, Ciro Beneduce, Bruno Lepri
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2507.00945v3 Announce Type: replace Abstract: Short-term forecasting of aggregated human mobility flows supports urban planning, intelligent transportation systems, and emergency response, yet existing models often require substantial mobility history and learn spatial structure implicitly thr...

📖 Read original article


338. Dynamic gain neuromodulation attenuates the stability gap under joint training ​

Author: Alejandro Rodriguez-Garcia, Anindya Ghosh, Srikanth Ramaswamy
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.NC

arXiv:2507.14056v3 Announce Type: replace Abstract: Recent work in continual learning has highlighted the stability gap -- a temporary performance drop on previously learned tasks when new ones are introduced. This phenomenon reflects a mismatch between rapid adaptation and strong retention at task ...

📖 Read original article


339. Physics-Informed Policy Iteration for High-Dimensional Hamilton--Jacobi--Bellman Equations: Interior Error Bounds without Boundary Data ​

Author: Yeongjong Kim, Minseok Kim, Yeoneung Kim, Namkyeong Cho
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.NA, math.NA

arXiv:2508.01718v2 Announce Type: replace Abstract: We develop a physics-informed policy-iteration method for stationary second-order Hamilton--Jacobi--Bellman equations arising in continuous-time stochastic control. Each policy-evaluation step is a linear elliptic PDE and is approximated by a mesh-...

📖 Read original article


340. Learning Multi-Timescale Interventions under Safety and Resource Constraints ​

Author: David Mguni, Wanrong Yang, Jing Dong, Ziquan Liu, Muhammad Salman Haleem, Baoxiang Wang, Dominik Wojtczak
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2508.03875v2 Announce Type: replace Abstract: Many sequential decision problems offer qualitatively different ways of influencing the environment: some interventions act immediately, whereas others induce persistent effects that continue to shape future states long after the decision that init...

📖 Read original article


341. Benchmarking the Robustness of Agentic Systems to Adversarially-Induced Harms ​

Author: Jonathan N"other, Adish Singla, Goran Radanovic
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2508.16481v3 Announce Type: replace Abstract: Ensuring the safe use of agentic systems requires a thorough understanding of the range of malicious behaviors these systems may exhibit. In this paper, we evaluate the robustness of LLM-based agentic systems against attacks that aim to elicit harm...

📖 Read original article


342. Deep Residual Echo State Networks: exploring residual orthogonal connections in untrained Recurrent Neural Networks ​

Author: Matteo Pinna, Andrea Ceni, Claudio Gallicchio
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2508.21172v3 Announce Type: replace Abstract: Echo State Networks (ESNs) are a particular type of untrained Recurrent Neural Networks (RNNs) within the Reservoir Computing (RC) framework, popular for their fast and efficient learning. However, traditional ESNs often struggle with long-term inf...

📖 Read original article


343. Score-based Membership Inference on Diffusion Models ​

Author: Mingxing Rao, Bowen Qu, Daniel Moyer
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2509.25003v3 Announce Type: replace Abstract: Membership inference attacks (MIAs) against Diffusion Models (DMs) raise pressing privacy concerns by revealing whether a sample was part of the training set. While existing methods typically rely on measuring reconstruction error across multiple d...

📖 Read original article


344. Self-Attention to Operator Learning-based 3D-IC Thermal Simulation ​

Author: Zhen Huang, Hong Wang, Wenkai Yang, Muxi Tang, Depeng Xie, Ting-Jung Lin, Yu Zhang, Wei W. Xing, Lei He
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR

arXiv:2510.15968v2 Announce Type: replace Abstract: Thermal management in 3D ICs is increasingly challenging due to higher power densities. Traditional PDE-solving-based methods, while accurate, are too slow for iterative design. Machine learning approaches like FNO provide faster alternatives but s...

📖 Read original article


345. NeuroAda: Activating Each Neuron's Potential for Parameter-Efficient Fine-Tuning ​

Author: Zhi Zhang, Yixian Shen, Congfeng Cao, Ekaterina Shutova
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2510.18940v2 Announce Type: replace Abstract: Existing parameter-efficient fine-tuning (PEFT) methods primarily fall into two categories: addition-based and selective in-situ adaptation. The former, such as LoRA, introduce additional modules to adapt the model to downstream tasks, offering str...

📖 Read original article


346. Rethinking Reasoning with MDLMs: Early Exits, Post-hoc Reasoning, and Beyond ​

Author: Zachary Horvitz, Raghav Singhal, Hao Zou, Carles Domingo-Enrich, Zhou Yu, Rajesh Ranganath, Kathleen McKeown
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.19990v2 Announce Type: replace Abstract: The reasoning paradigm, where language models reason before answering, has enabled breakthroughs on tasks such as mathematical problem-solving. While current tooling for reasoning is built around next-token prediction trained models, recent works i...

📖 Read original article


347. flowengineR: A Modular and Extensible Framework for Fair and Reproducible Workflow Design in R ​

Author: Maximilian Willer, Peter Ruckdeschel
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CY, stat.ME

arXiv:2511.00079v2 Announce Type: replace Abstract: flowengineR is an R package designed to provide a modular and extensible framework for building reproducible algorithmic workflows for general-purpose machine learning pipelines. It is motivated by the rapidly evolving field of algorithmic fairness...

📖 Read original article


348. Directional-Clamp PPO ​

Author: Gilad Karpel, Ruida Zhou, Shoham Sabach, Mohammad Ghavamzadeh
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.02577v2 Announce Type: replace Abstract: Proximal Policy Optimization (PPO) is widely regarded as one of the most successful deep reinforcement learning algorithms, known for its robustness and effectiveness across a range of problems. The PPO objective encourages the importance ratio bet...

📖 Read original article


349. iLTM: Integrated Large Tabular Model ​

Author: David Bonet, Mar\c{c}al Comajoan Cara, Alvaro Calafell, Daniel Mas Montserrat, Alexander G. Ioannidis
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2511.15941v2 Announce Type: replace Abstract: Tabular data underpins decisions across science, industry, and public services. Despite rapid progress, advances in deep learning have not fully carried over to the tabular domain, where gradient-boosted decision trees (GBDTs) remain a default choi...

📖 Read original article


350. Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM ​

Author: Adarsh Kumarappan, Ayushi Mehrotra
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2511.18721v4 Announce Type: replace Abstract: The SmoothLLM defense provides a certification guarantee against jailbreaking attacks, but it relies on a strict "k-unstable" assumption that rarely holds in practice. This strong assumption can limit the trustworthiness of the provided safety cert...

📖 Read original article


351. Automating Deception: Scalable Multi-Turn LLM Jailbreaks ​

Author: Adarsh Kumarappan, Ananya Mujoo
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2511.19517v3 Announce Type: replace Abstract: Multi-turn conversational attacks, which leverage psychological principles like Foot-in-the-Door (FITD), where a small initial request paves the way for a more significant one, to bypass safety alignments, pose a persistent threat to Large Language...

📖 Read original article


352. Benchmarking In-context Experiential Learning Through Repeated Product Recommendations ​

Author: Gilbert Yang, Yaqin Chen, Thomson Yen, Hongseok Namkoong
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.22130v2 Announce Type: replace Abstract: To navigate ever-shifting real-world environments, agents must grapple with incomplete knowledge and adapt their strategies through experience. However, current evaluations of LLM-based agents largely overlook this capability. Crucially, we stress ...

📖 Read original article


353. Mitigating Barren Plateaus in Quantum Denoising Diffusion Probabilistic Model ​

Author: Haipeng Cao, Kaining Zhang, Dacheng Tao, Zhaofeng Su
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, quant-ph

arXiv:2512.06695v3 Announce Type: replace Abstract: Quantum generative models exploit quantum superposition and entanglement to enhance learning efficiency for both classical and quantum data. Recently, inspired by classical diffusion frameworks, the quantum denoising diffusion probabilistic model h...

📖 Read original article


354. Transformers for Multimodal Brain State Decoding: Integrating Functional Magnetic Resonance Imaging Data and Medical Metadata ​

Author: Danial Jafarzadeh Jazi, Maryam Hajiesmaeili
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.08462v2 Announce Type: replace Abstract: Decoding brain states from functional magnetic resonance imaging (fMRI) data is vital for advancing neuroscience and clinical applications. While traditional machine learning and deep learning approaches have made strides in leveraging the high-dim...

📖 Read original article


355. Fundamental Limitations of Favorable Privacy-Utility Guarantees for DP-SGD ​

Author: Murat Bilgehan Ertan, Marten van Dijk
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2601.10237v3 Announce Type: replace Abstract: Differentially Private Stochastic Gradient Descent (DP-SGD) is the dominant paradigm for private training, but its fundamental limitations under worst-case adversarial privacy definitions remain poorly understood. We analyze DP-SGD in the $f$-diffe...

📖 Read original article


356. Hybrid Mamba-Attention Neural Architecture for Channel Estimation ​

Author: Dianxin Luan, Chengsi Liang, Jie Huang, Zheng Lin, Kaitao Meng, John Thompson, Cheng-Xiang Wang, Ozgur Akan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP

arXiv:2601.17108v3 Announce Type: replace Abstract: This paper proposes a hybrid Mamba-attention neural architecture to achieve improved channel estimation for orthogonal frequency-division multiplexing (OFDM) waveforms, particularly for configurations with a large number of subcarriers. By integrat...

📖 Read original article


357. OATS: Online Data Augmentation for Time Series Foundation Models ​

Author: Junwei Deng, Chang Xu, Jiaqi W. Ma, Ming Jin, Chenghao Liu, Xu Zhang, Li Zhao, Jiang Bian
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.19040v2 Announce Type: replace Abstract: Time Series Foundation Models (TSFMs) are a powerful paradigm for time series analysis and are often enhanced by synthetic data augmentation to improve the training data quality. Existing augmentation methods, however, typically rely on heuristics ...

📖 Read original article


358. Temporal Sepsis Modeling: a Relational and Explainable-by-Design Framework ​

Author: Vincent Lemaire, N'edra Meloulli, Pierre Jaquet
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.21747v4 Announce Type: replace Abstract: Sepsis remains one of the most complex and heterogeneous syndromes in intensive care. While deep learning models achieve competitive performance in early sepsis prediction, their decision processes often remain difficult to interpret clinically, an...

📖 Read original article


359. Shattered Compositionality: Counterintuitive Learning Dynamics of Transformers for Arithmetic ​

Author: Xingyu Zhao, Darsh Sharma, Rheeya Uppaal, Yiqiao Zhong
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.22510v2 Announce Type: replace Abstract: Large language models (LLMs) often achieve strong benchmark accuracy yet remain brittle under small distribution shifts. While recent mechanistic studies reveal the discrepancy between LLMs and humans in skill compositions, the learning dynamics of...

📖 Read original article


360. OD-Gear: Online Decomposition and Group Sampling for Expert-Guided Adversarial Routing in Scalable Capacitated Vehicle Routing ​

Author: Dongbin Jiao, Zisheng Chen, Xianyi Wang, Jintao Shi, Shengcai Liu, Shi Yan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.00488v3 Announce Type: replace Abstract: Solving large-scale capacitated vehicle routing problems (CVRP) is hindered by the high complexity of classical heuristics and the limited generalization of neural solvers. To bridge this gap, we propose OD-Gear, an expert-guided adversarial framew...

📖 Read original article


361. Beyond the Node: Clade-level Selection for Efficient MCTS in Automatic Heuristic Design ​

Author: Kezhao Lai, Yutao Lai, Hai-Lin Liu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.00549v2 Announce Type: replace Abstract: While Monte Carlo Tree Search (MCTS) shows promise in Large Language Model (LLM) based Automatic Heuristic Design (AHD), it suffers from a critical over-exploitation tendency under the limited computational budgets required for heuristic evaluation...

📖 Read original article


362. Test-time Generalization for Physics through Neural Operator Splitting ​

Author: Louis Serrano, Jiequn Han, Edouard Oyallon, Shirley Ho, Rudy Morel
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.00884v2 Announce Type: replace Abstract: Neural operators have shown promise in learning solution maps of partial differential equations (PDEs), but they often struggle to generalize when test inputs lie outside the training distribution, such as novel initial conditions, unseen PDE coeff...

📖 Read original article


363. Diving into Kronecker Adapters: Component Design Matters ​

Author: Jiayu Bai, Danchen Yu, Zhenyu Liao, TianQi Hou, Feng Zhou, Robert C. Qiu, Zenan Ling
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.01267v3 Announce Type: replace Abstract: Kronecker adapters have emerged as a promising approach for fine-tuning large-scale models, enabling high-rank updates through tunable component structures. However, existing work largely treats the component structure as a fixed or heuristic desig...

📖 Read original article


364. Universal One-third Time Scaling in Learning Peaked Distributions ​

Author: Yizhou Liu, Ziming Liu, Cengiz Pehlevan, Jeff Gore
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2602.03685v3 Announce Type: replace Abstract: Training large language models (LLMs) is computationally expensive, partly because the loss exhibits slow power-law convergence whose origin remains debatable. Through systematic analysis of toy models and empirical evaluation of LLMs, we show that...

📖 Read original article


365. On the Infinite Width and Depth Limits of Predictive Coding Networks ​

Author: Francesco Innocenti, El Mehdi Achour, Rafal Bogacz
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE

arXiv:2602.07697v3 Announce Type: replace Abstract: Predictive coding (PC) is a biologically plausible alternative to standard backpropagation (BP) that minimises an energy function with respect to network activities before updating weights. Recent work has improved the training stability of deep PC...

📖 Read original article


366. What Does Preference Learning Recover from Pairwise Comparison Data? ​

Author: Rattana Pukdee, Maria-Florina Balcan, Pradeep Ravikumar
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.10286v3 Announce Type: replace Abstract: Pairwise preference learning is central to machine learning, with recent applications in aligning language models with human preferences. A typical dataset consists of triplets $(x, y^+, y^-)$, where response $y^+$ is preferred over response $y^-$ ...

📖 Read original article


367. SMAC: Score-Matched Actor-Critics for Robust Offline-to-Online Transfer ​

Author: Nathan Samuel de Lara, Florian Shkurti
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17632v3 Announce Type: replace Abstract: Modern offline Reinforcement Learning (RL) methods find performant actor-critics, however, fine-tuning these actor-critics online with value-based RL algorithms typically causes immediate drops in performance. We provide evidence consistent with th...

📖 Read original article


368. Doubly Adaptive Channel and Spatial Attention for Semantic Image Communication by IoT Devices ​

Author: Soroosh Miri, Sepehr Abolhasani, Shahrokh Farahmand, S. Mohammad Razavizadeh, Jiguang He
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.22794v2 Announce Type: replace Abstract: Internet of Things (IoT) networks face significant challenges such as limited communication bandwidth, constrained computational and energy resources, and highly dynamic wireless channel conditions. Utilization of deep neural networks (DNNs) combin...

📖 Read original article


369. Attn-QAT: 4-Bit Attention With Quantization-Aware Training ​

Author: Peiyuan Zhang, Matthew Noto, Wenxuan Tan, Chengquan Jiang, Will Lin, Wei Zhou, Hao Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.00040v3 Announce Type: replace Abstract: Achieving reliable 4-bit attention is a prerequisite for end-to-end FP4 computation on emerging FP4-capable GPUs, yet attention remains the main obstacle due to FP4's tiny dynamic range and attention's heavy-tailed activations. This paper presents ...

📖 Read original article


370. Online Learnability of Chain-of-Thought Verifiers: Soundness and Completeness Trade-offs ​

Author: Maria-Florina Balcan, Avrim Blum, Kiriaki Fragkia, Zhiyuan Li, Dravyansh Sharma
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.03538v4 Announce Type: replace Abstract: Large Language Models (LLMs) using chain-of-thought reasoning have demonstrated great potential for solving complex reasoning and planning tasks. However, their outputs remain unreliable and need careful verification. Even as LLMs get more accurate...

📖 Read original article


371. Adversarial Latent-State Training for Robust Policies in Partially Observable Domains ​

Author: Angad Singh Ahuja
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2603.07313v4 Announce Type: replace Abstract: Robustness under latent distribution shift remains challenging in partially observable reinforcement learning. We formalize a focused setting where an adversary selects a hidden initial latent distribution before the episode, termed an adversarial ...

📖 Read original article


372. Estimating Condition Number with Graph Neural Networks ​

Author: Erin Carson, Xinye Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2603.10277v5 Announce Type: replace Abstract: In this paper, we propose a fast method for estimating the condition number of sparse matrices using graph neural networks (GNNs). For efficient deployment of GNNs, we introduce a graph feature construction with $\mathrm{O}(\mathrm{nnz} + n)$ compl...

📖 Read original article


373. Quantifying Membership Disclosure Risk for Tabular Synthetic Data Using Kernel Density Estimators ​

Author: Rajdeep Pathak, Amit Basak, Sayantee Jana
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.AP

arXiv:2603.10937v2 Announce Type: replace Abstract: The use of synthetic data has become increasingly popular as a privacy-preserving alternative to sharing real datasets, especially in sensitive domains such as healthcare, finance, and demography. However, the privacy assurances of synthetic data a...

📖 Read original article


374. When should we trust the annotation? Selective prediction for molecular structure retrieval from mass spectra ​

Author: Mira J"urgens, Gaetan De Waele, Morteza Rakhshaninejad, Willem Waegeman
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2603.10950v2 Announce Type: replace Abstract: Machine learning methods for identifying molecular structures from tandem mass spectra (MS/MS) have advanced rapidly, yet current approaches still exhibit significant error rates. In high-stakes applications such as clinical metabolomics and enviro...

📖 Read original article


375. ROM: Real-time Overthinking Mitigation via Streaming Detection and Intervention ​

Author: Xinyan Wang, Xiaogeng Liu, Ming Pei, Chaowei Xiao
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2603.22016v3 Announce Type: replace Abstract: Large Reasoning Models (LRMs) often reach a correct solution before their long Chain-of-Thought trace ends, yet continue with redundant verification, repeated attempts, or unnecessary exploration that wastes computation and can even overturn the co...

📖 Read original article


376. SPA: A Simple but Tough-to-Beat Baseline for Knowledge Injection ​

Author: Kexian Tang, Jiani Wang, Shaowen Wang, Kaifeng Lyu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2603.22213v2 Announce Type: replace Abstract: While large language models (LLMs) are pretrained on massive amounts of data, their knowledge coverage remains incomplete in specialized, data-scarce domains, motivating extensive efforts to study synthetic data generation for knowledge injection. ...

📖 Read original article


377. A Sobering Look at Tabular Data Generation via Probabilistic Circuits ​

Author: Davide Scassola, Dylan Ponsford, Adri'an Javaloy, Sebastiano Saccani, Luca Bortolussi, Henry Gouk, Antonio Vergari
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.23016v2 Announce Type: replace Abstract: Tabular data is more challenging to generate than text and images, due to its heterogeneous features and much lower sample sizes. On this task, diffusion-based models are the current state-of-the-art (SotA) model class, achieving almost perfect per...

📖 Read original article


378. Explaining, Verifying, and Aligning Semantic Hierarchies in Vision-Language Model Embeddings ​

Author: Gesina Schwalbe, Mert Keser, Moritz Bayerkuhnlein, Edgar Heinert, Annika M"utze, Marvin Keller, Sparsh Tiwari, Georgii Mikriukov, Diedrich Wolter, Jae Hee Lee, Matthias Rottmann
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.26798v2 Announce Type: replace Abstract: Vision-language model (VLM) encoders such as CLIP enable strong retrieval and zero-shot classification in a shared image-text embedding space, yet the semantic organization of this space is rarely inspected. We present a post-hoc framework to expla...

📖 Read original article


379. From Independent to Correlated Diffusion: Generalized Generative Modeling with Probabilistic Computers ​

Author: Nihal Sanjay Singh, Mazdak Mohseni-Rajaee, Shaila Niazi, Kerem Y. Camsari
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.ET

arXiv:2603.27996v2 Announce Type: replace Abstract: Diffusion models have emerged as a powerful framework for generative tasks in deep learning. They decompose generative modeling into two computational primitives: deterministic neural-network evaluation and stochastic sampling. Current implementati...

📖 Read original article


380. Critic-Free Deep Reinforcement Learning for Maritime Coverage Path Planning on Irregular Hexagonal Grids ​

Author: Carlos S. Sep'ulveda, Gonzalo A. Ruz
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE, cs.RO

arXiv:2603.28385v2 Announce Type: replace Abstract: Maritime surveillance missions, such as search and rescue and environmental monitoring, rely on the efficient allocation of sensing assets over vast and geometrically complex areas. Traditional Coverage Path Planning (CPP) approaches depend on deco...

📖 Read original article


381. From Rebound to Remedy: Understanding and Mitigating Reward Hacking via Representation Engineering ​

Author: Rui Wu, Ruixiang Tang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2604.01476v2 Announce Type: replace Abstract: Reinforcement learning for LLMs is vulnerable to reward hacking, where models exploit shortcuts to maximize reward without solving the intended task. We systematically study this phenomenon in coding tasks using an environment-manipulation setting,...

📖 Read original article


382. Matching Accuracy, Different Geometry: Evolution Strategies vs GRPO in LLM Post-Training ​

Author: William Hoy, Binxu Wang, Xu Pan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.01499v2 Announce Type: replace Abstract: Evolution Strategies (ES) have emerged as a scalable gradient-free alternative to reinforcement learning based LLM fine-tuning, but it remains unclear whether comparable task performance implies comparable solutions in parameter space. We compare E...

📖 Read original article


383. Cognitive Energy Modeling for Neuroadaptive Human-Machine Systems using EEG and WGAN-GP ​

Author: Sriram Sattiraju, Vaibhav Gollapalli, Aryan Shah, Timothy McMahan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.HC

arXiv:2604.01653v2 Announce Type: replace Abstract: Electroencephalography (EEG) provides a non-invasive insight into the brain's cognitive and emotional dynamics. However, modeling how these states evolve in real time and quantifying the energy required for such transitions remains a major challeng...

📖 Read original article


384. Are Latent Reasoning Models Easily Interpretable? ​

Author: Connor Dilgren, Sarah Wiegreffe
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.04902v2 Announce Type: replace Abstract: Latent reasoning models (LRMs) have attracted significant research interest due to their low inference cost (relative to explicit reasoning models) and theoretical ability to explore multiple reasoning paths in parallel. However, these benefits com...

📖 Read original article


385. Not All Turns Are Equally Hard: Adaptive Thinking Budgets For Efficient Multi-Turn Reasoning in Agents ​

Author: Neharika Jali, Anupam Nayak, Gauri Joshi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.05164v3 Announce Type: replace Abstract: As LLM reasoning performance plateaus, improving inference-time compute efficiency is crucial to mitigate overthinking and long thinking traces even for simple queries. Prior approaches including length regularization, adaptive routing, and difficu...

📖 Read original article


386. Weighted Bayesian Conformal Prediction ​

Author: Xiayin Lou, Peng Luo
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, physics.app-ph, stat.ML

arXiv:2604.06464v3 Announce Type: replace Abstract: Machine learning predictors are rarely deployed on data that match the data they were calibrated on. Conformal prediction and its risk-control generalization promise distribution-free guarantees at deployment, but only under exchangeability or a co...

📖 Read original article


387. Preference Redirection via Attention Concentration: An Attack on Computer Use Agents ​

Author: Dominik Seip, Matthias Hein
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.08005v2 Announce Type: replace Abstract: Advancements in multimodal foundation models have enabled the development of Computer Use Agents (CUAs) capable of autonomously interacting with GUI environments. As CUAs are not restricted to certain tools, they allow to automate more complex agen...

📖 Read original article


388. Adalina: Adaptive Linear Approximation for the Shapley Value and Beyond ​

Author: Weida Li, Yaoliang Yu, Bryan Kian Hsiang Low
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.08438v3 Announce Type: replace Abstract: The Shapley value, and its broader family of semi-values, has received much attention in various attribution problems. A fundamental and long-standing challenge is their efficient approximation, since exact computation generally requires an exponen...

📖 Read original article


389. In-context superposition: human-like working memory interference in large language models ​

Author: Hua-Dong Xiong, Li Ji-An, Jiaqi Huang, Robert C. Wilson, Kwonjoon Lee, Xue-Xin Wei
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.09670v2 Announce Type: replace Abstract: Intelligent systems must maintain and manipulate task-relevant information online to adapt to dynamic environments and changing goals. This capacity, known as working memory, is fundamental to human reasoning and intelligence. Despite their radical...

📖 Read original article


390. Symmetry Reveals Layerwise Dynamics: How Transformers Perform In-Context Classification ​

Author: Patrick Lutz, Themistoklis Haris, Arjun Chandra, Aditya Gangrade, Venkatesh Saligrama
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.11613v4 Announce Type: replace Abstract: Transformers can perform in-context classification from a few labeled examples, yet the inference-time algorithm remains opaque. We study multi-class linear classification in the hard no-margin regime and make the computation identifiable by enforc...

📖 Read original article


391. Fairness is Not Flat: Geometric Phase Transitions Against Shortcut Learning ​

Author: Nicolas Rodriguez-Alvarez (Instituto de Educacion Secundaria Parquesol, Valladolid, Spain)
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.11704v2 Announce Type: replace Abstract: Deep Neural Networks are highly susceptible to shortcut learning, frequently memorizing low-dimensional spurious correlations instead of underlying causal mechanisms. This phenomenon not only degrades out-of-distribution robustness but also induces...

📖 Read original article


392. Multi-Head Residual-Gated DeepONet for Coherent Nonlinear Wave Dynamics ​

Author: Zhiwei Fan, Yiming Pan, Daniel Coca
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.11972v2 Announce Type: replace Abstract: Coherent nonlinear wave dynamics are often strongly shaped by a compact set of physically meaningful descriptors of the initial state. Traditional neural operators typically treat the input-output mapping as a largely black-box high-dimensional reg...

📖 Read original article


393. Hybrid Quantum-Classical PINNs for Scientific Computing: A Multi-GPU Open-Source Framework ​

Author: Ziv Chen, Hemanth Chandravamsi, Shimon Pisnoy, Aaron Goldgewert, Gal Shaviner, Boris Shragner, Steven H. Frankel
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph, quant-ph

arXiv:2604.15645v2 Announce Type: replace Abstract: We present QPINNACLE, an open-source computational framework for physics-informed neural networks (PINNs) that integrates modern training strategies, multi-GPU acceleration, and hybrid quantum-classical architectures within a unified modular workfl...

📖 Read original article


394. Label-Efficient Bilateral Attention for Parkinson's Disease Screening from Wrist-Worn IMU Signals ​

Author: Meheru Zannat
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.18372v2 Announce Type: replace Abstract: Parkinson's disease (PD) is a chronic neurodegenerative disorder. It shows multiple motor symptoms such as tremor, bradykinesia, postural instability, and freezing of gait (FoG). PD is currently diagnosed clinically through physical examination by ...

📖 Read original article


395. Switching Theory for Q-Learning ​

Author: Donghwan Lee
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY

arXiv:2604.19569v5 Announce Type: replace Abstract: Q-learning is a fundamental algorithmic primitive in reinforcement learning. This paper develops a new framework for analyzing constant step-size tabular Q-learning from a switching linear system (SLS) viewpoint. In particular, we derive a stochast...

📖 Read original article


396. From Local to Cluster: A Unified Framework for Causal Discovery with Latent Variables ​

Author: Zongyu Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.22416v3 Announce Type: replace Abstract: Latent variables pose a fundamental obstacle to both causal discovery and inference. Local approaches exploiting direct neighborhood relations provide little beyond immediate dependencies. Cluster-level methods, though capable of broader reasoning,...

📖 Read original article


397. Process Supervision of Confidence Margin for Calibrated LLM Reasoning ​

Author: Liaoyaqi Wang, Chunsheng Zuo, William Jurayj, Benjamin Van Durme, Anqi Liu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2604.23333v2 Announce Type: replace Abstract: Scaling test-time computation with reinforcement learning (RL) has emerged as a reliable path to improve large language models (LLM) reasoning ability. Yet, outcome-based reward often incentivizes models to be overconfident, leading to hallucinatio...

📖 Read original article


398. Dynamic Regret for Online Regression in RKHS via Discounted VAW and Subspace Approximation ​

Author: Dmitry B. Rokhlin, Georgiy A. Karapetyants
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.25021v2 Announce Type: replace Abstract: We study online regression with the square loss in a reproducing kernel Hilbert space under a dynamic regret criterion. The learner is compared with a time-varying comparator sequence, and the bounds depend on its path length in the RKHS norm. The ...

📖 Read original article


399. Barriers to Universal Reasoning With Transformers (And How to Overcome Them) ​

Author: Oliver Kraus, Yash Sarrof, Yuekun Yao, Alexander Koller, Michael Hahn
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2604.25800v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) has been shown to empirically improve Transformers' performance, and theoretically increase their expressivity to Turing completeness. However, whether Transformers can learn to generalize to CoT traces longer than those seen...

📖 Read original article


400. Rethinking KV Cache Eviction via a Unified Information-Theoretic Objective ​

Author: Jiaming Yang, Chenwei Tang, Liangli Zhen, Jiancheng Lv
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT

arXiv:2604.25975v2 Announce Type: replace Abstract: Key-Value (KV) caching is essential for large language model inference, yet its memory overhead poses a critical bottleneck for long-context generation. Existing eviction policies predominantly rely on empirical heuristics, lacking a rigorous theor...

📖 Read original article


401. Uncertainty-Aware Predictive Safety Filters for Probabilistic Neural Network Dynamics ​

Author: Bernd Frauenknecht, Lukas Kesper, Daniel Mayfrank, Henrik Hose, Sebastian Trimpe
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2604.26836v3 Announce Type: replace Abstract: Predictive safety filters (PSFs) leverage model predictive control to enforce constraint satisfaction during deep reinforcement learning (RL) exploration, yet their reliance on first-principles models or Gaussian processes limits scalability and br...

📖 Read original article


402. Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback ​

Author: Yikai Wang, Shang Liu, Jose Blanchet
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, math.OC, stat.ML

arXiv:2605.00155v4 Announce Type: replace Abstract: Reinforcement learning from human feedback (RLHF) is a central post-training tool for aligning large language models, but its training reward is only a learned proxy for true human utility. This creates a decision problem under objective misspecifi...

📖 Read original article


403. Uncertainty in a Single Pass: A Closed-Form Identity for One-Step Flow Matching ​

Author: Jiarui Xing, Song Wang, Jian Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2605.00941v5 Announce Type: replace Abstract: Flow matching provides a highly effective framework for generative modeling, yet estimating the uncertainty of its generated samples remains a fundamental challenge. Existing methods rely on auxiliary variance heads, model ensembles, or iterative c...

📖 Read original article


404. Stable GFlowNets with TV Monitoring and Probabilistic Guarantees ​

Author: Zengxiang Lei, Ananth Shreekumar, Jonathan Rosenthal, Ruoyu Song, Alvaro A. Cardenas, Daniel J. Fremont, Dongyan Xu, Satish Ukkusuri, Z. Berkay Celik
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2605.01729v3 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) sample diverse structured objects in proportion to reward and have been applied to molecular discovery and biological-sequence design, where finding multiple high-quality candidates is more useful than returning...

📖 Read original article


405. Statistically-Lossless Quantization of Large Language Models ​

Author: Michael Helcig, Eldar Kurtic, Dan Alistarh
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.02404v2 Announce Type: replace Abstract: Model quantization has become essential for efficient large language model deployment, yet existing approaches present clear trade-offs: methods such as GPTQ and AWQ achieve practical compression but are lossy, while lossless techniques preserve fi...

📖 Read original article


406. A Coupled Physics-Informed Neural Network for Greenhouse Climate State Reconstruction and Parameter Identification under Sparse Sensor Measurements ​

Author: Sani Biswas, Khursheed J. Ansari, Md. Nasim Akhtar
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.02524v2 Announce Type: replace Abstract: Accurate reconstruction of greenhouse climate variables from sparse sensor measurements is essential for intelligent environmental monitoring, automated climate control, and precision agriculture. In practical greenhouse operation, sensor failures,...

📖 Read original article


407. Aggregation in conformal e-classification ​

Author: Vladimir Vovk
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.07963v2 Announce Type: replace Abstract: Aggregating conformal predictors is a standard way of balancing their predictive and computational efficiency while retaining their validity, at least approximately. An important advantage of conformal e-predictors is that they are easier to aggreg...

📖 Read original article


408. The Safety-Aware Denoiser for Text Diffusion Models ​

Author: Amman Yusuf, Zhejun Jiang, Mijung Park
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.08116v3 Announce Type: replace Abstract: Recent work on text diffusion models offers a promising alternative to autoregressive generation, but controlling their safety remains underexplored. Existing safety approaches are geared toward autoregressive models and typically rely on post-hoc ...

📖 Read original article


409. Early Data Exposure Improves Robustness to Subsequent Fine-Tuning ​

Author: Lawrence Feng, Gaurav R. Ghosal, Jacob Mitchell Springer, Ziqian Zhong, Aditi Raghunathan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.12705v2 Announce Type: replace Abstract: How can we train models whose post-trained capabilities survive subsequent fine-tuning? Rather than focusing on downstream interventions to mitigate forgetting of upstream capabilities, we study how upstream training choices - that is, the manner i...

📖 Read original article


410. Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy ​

Author: Adarsh Kumarappan, Ananya Mujoo
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.12991v3 Announce Type: replace Abstract: LLM-based multi-agent pipelines flip from correct to incorrect answers under simulated peer disagreement at rates we term yield, a vulnerability widely attributed to RLHF-induced sycophancy. We test this attribution across four model families and f...

📖 Read original article


411. DeltaPrompts: Escaping the Zero-Delta Trap in Multimodal Distillation ​

Author: Jaehun Jung, Hyunwoo Kim, Brandon Cui, Ximing Lu, David Acuna, Prithviraj Ammanabrolu, Yejin Choi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2605.15532v3 Announce Type: replace Abstract: Distillation enables compact Vision-Language Models (VLMs) to obtain strong reasoning capabilities, yet the prompts driving this process are typically chosen via simple heuristics or aggregated from off-the-shelf datasets. We reveal a critical inef...

📖 Read original article


412. ArtifactLinker: Linking Scientific Artifacts for Automatic State-of-the-Art Discovery ​

Author: Haofei Yu, Jiaxuan You, Peter Clark, Bodhisattwa Prasad Majumder, Kyle Richardson
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.16902v2 Announce Type: replace Abstract: Scientific artifacts such as models and datasets are foundations for research. With the rapid growth of platforms like HuggingFace, researchers now have access to a large number of artifacts. Yet, a key challenge remains: how can we automatically d...

📖 Read original article


413. Efficient and Noise-Tolerant PAC Learning of Multiclass Linear Classifiers ​

Author: Rita Adhikari, Shiwei Zeng
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.18662v2 Announce Type: replace Abstract: Noise-tolerant PAC learning of linear models has been of central interests in machine learning community since the last century. In recent years, many computationally-efficient algorithms have been proposed for the problem of learning linear thresh...

📖 Read original article


414. Dimensional Balance Improves Large Scale Spatiotemporal Prediction Performance ​

Author: Jing Chen, Shixiang Pan, Yujie Fan, Haocheng Ye, Haitao Xu, Wenqiang Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.18793v2 Announce Type: replace Abstract: Accurate spatiotemporal pattern analysis is critical in fields such as urban traffic, meteorology, and public health monitoring. However, existing methods face performance bottlenecks, typically yielding only incremental gains and often exhibiting ...

📖 Read original article


415. Robust Recommendation from Noisy Implicit Feedback: A GMM-Weighted Bayes-label Transition Matrix Framework ​

Author: Zongyu Li, Xuanyu Liu, Gongce Cao, Shirui Sun, Yaqi Fang, Yongshuai Yu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.20721v2 Announce Type: replace Abstract: Label noise is a central challenge in learning from implicit feedback for recommendation. Conventional approaches discard noisy examples for robustness, but this sacrifices data efficiency. Unlike filtering approaches, Bayes-label transition matrix...

📖 Read original article


416. How Many Different Outputs Can a Transformer Generate? ​

Author: Maxime Meyer, Mario Michelessa, Caroline Chaux, Vincent Y. F. Tan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.22223v2 Announce Type: replace Abstract: We study how we can leverage only a handful of characteristics of a transformer's architecture to closely predict the number of different sequences it can output, both qualitatively and quantitatively. We provide an upper bound depending on the len...

📖 Read original article


417. Latent Spectroscopy: Posterior Collapse as a Feature ​

Author: Johannes Hirn
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech

arXiv:2605.22691v2 Announce Type: replace Abstract: We show that, in linear Gaussian VAEs, posterior collapse is a form of latent feature selection. Feature importance is set by each latent coordinate's contribution to reconstruction, itself given by the corresponding PCA eigenvalue. Varying the reg...

📖 Read original article


418. Don't Retrain, Just Reuse: Recovering Dual-Target Molecules from Single-Target Diffusion Models ​

Author: Qingyuan Zeng, Pengxiang Cai, Zixin Guan, Ziyang Chen, Anglin Liu, Xinyao Lai, Jintai Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.25681v2 Announce Type: replace Abstract: Designing a single molecule that modulates two targets is a promising strategy for polypharmacology, but it remains substantially harder than standard single-target generation because one candidate must satisfy two binding requirements while preser...

📖 Read original article


419. Test-Time Collective Action: Proxy-Based Perturbations for Correcting Algorithmic Harms ​

Author: Meghana Bhange, Ulrich A"ivodji, Elliot Creager
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2605.27689v2 Announce Type: replace Abstract: When machine learning systems under-perform for particular subgroups, affected users typically have no way to correct these disparities without relying on platform-level fixes. Existing approaches to algorithmic fairness rely on provider-centric ap...

📖 Read original article


420. Convex Basins in Single-Index Model Loss Landscapes: Applications to Robust Recovery under Strong Adversarial Corruption ​

Author: Santanu Das, Sagnik Chatterjee, Jatin Batra
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.29497v2 Announce Type: replace Abstract: We study the problem of robustly learning Gaussian Single Index Models (SIMs) in the presence of heavy-tailed noise and a constant fraction of adversarially corrupted covariates and responses. Prior work on robust recovery has considered settings s...

📖 Read original article


421. Lagrangian Perturbation Diffusion Steering: Latent Reinforcement Learning for Generative Policies ​

Author: Hikmet Simsir, Ozgur S. Oguz
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.01151v2 Announce Type: replace Abstract: Behavior cloning with high-capacity generative policies achieves strong imitation performance, but is often limited by demonstration coverage and distribution shift. Direct reinforcement learning fine-tuning can improve performance, but updating la...

📖 Read original article


422. Minimax-Optimal Policy Regret in Partially Observable Markov Games ​

Author: Raman Arora
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2606.02363v2 Announce Type: replace Abstract: We study sequential decision-making in partially observable environments against strategic, adaptive opponents, modeled as partially observable Markov games (POMGs). The central challenge is to learn latent dynamics from partial observations while ...

📖 Read original article


423. Skip a Layer or Loop It? Learning Program-of-Layers in LLMs ​

Author: Ziyue Li, Yang Li, Tianyi Zhou
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.06574v2 Announce Type: replace Abstract: Large language models (LLMs) perform inference by following a fixed depth and order, non-recurrent execution of all layers. We reveal the wide existence of training-free, flexible, dynamic program-of-layers (PoLar), where pretrained layers can be p...

📖 Read original article


424. Enhancing AI Interpretability with Localised Architectures ​

Author: Ian Seet, Jonas Bozenhard, Simon Ostermann
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.07998v3 Announce Type: replace Abstract: Recent advances in generative AI, especially powerful Large Language Models (LLMs), raise concerns over the interpretability, safety and sustainability of these large and opaque AI models. The power of such architectures is derived not only from th...

📖 Read original article


425. Alignment Collapse Under KV Cache Quantization: Diagnosis and Mitigation ​

Author: Bruce Changlong Xu, Adarsh Kumarappan, Mu Zhou
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET

arXiv:2606.09864v2 Announce Type: replace Abstract: Key-value (KV) cache quantization is widely used to reduce Large Language Model (LLM) inference memory, yet existing evaluations solely focus on measuring perplexity and accuracy without assessing the safety impact. In this study, we explore alignm...

📖 Read original article


426. TRAPS: Treatment-Assignment Prediction via Pathway-informed Stratification ​

Author: Sujoy Banik, Sayantan Chakraborty, Boishakhi Das Toma, Zainab Ghafoor, Ushashi Bhattacharjee, Koushik Howlader, Tirtho Roy
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.MA, q-bio.QM

arXiv:2606.09898v2 Announce Type: replace Abstract: Cancer treatment involves decisions across multiple clinical outcomes, yet pathway-informed deep learning models are typically evaluated in isolation, making their relative benefits unclear. We present a harmonized benchmark of three biologically i...

📖 Read original article


427. Online Shift Detection and Conformal Adaptation for Deployed Safety Classifiers ​

Author: Jun Wen Leong
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, stat.ML

arXiv:2606.11949v4 Announce Type: replace Abstract: Reasoning models deployed as safety monitors exhibit a systematic vulnerability: reasoning-token budget starvation. Adversarial inputs require $3.3\times$ more reasoning tokens than benign inputs to produce valid safety scores ($T_{50,\text{adv}}{=...

📖 Read original article


428. Two-Layer Linear Auto-Regressive Models Estimate Latent States ​

Author: Yahya Sattar, Sunmook Choi, Leo Maynard-Zhang, Yassir Jedra, Maryam Fazel, Sarah Dean
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY, math.OC, stat.ML

arXiv:2606.12691v2 Announce Type: replace Abstract: Auto-regressive models have emerged as powerful tools for sequential data, from language to video. Understanding how and why these models learn latent representations remains an open theoretical question. In this work, we demonstrate that when trai...

📖 Read original article


429. Placing Degree Scales After LayerNorm ​

Author: Yash Vardhan Tomar, Aryav Das
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.14022v3 Announce Type: replace Abstract: Graph neural networks (GNNs) are widely used to learn node-selection policies on graphs, and most stack graph attention (GAT) blocks with LayerNorm. On degree-sensitive tasks, LayerNorm tends to remove the degree signal these models need to rank no...

📖 Read original article


430. Brownian Kernel Ladders ​

Author: Mahdi Mohammadigohari, Giuseppe Di Fatta, Giuseppe Nicosia, Panos M Pardalos
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.15812v2 Announce Type: replace Abstract: We introduce Brownian kernel ladders (BKLs), a recursive hierarchy of integral reproducing kernel Hilbert spaces built from linear functionals by repeatedly integrating Brownian pullback kernels indexed by functions from the preceding layer. The no...

📖 Read original article


431. Learning aligned EEG representations with subject-specific encoders ​

Author: Bruna J. Lopes, Gabriel Schwartz, Sylvain Chevallier, Raphael Y. de Camargo, Bruno Aristimunha
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.16462v2 Announce Type: replace Abstract: Cross-subject EEG decoding promises more training data, but it also exposes neural networks to strong inter-subject distribution shifts. We study whether task supervision and architecture alone can learn subject-aligned representations. We replace ...

📖 Read original article


432. Unsupervised Disentanglement Without Compromises : How Functional Orthogonality Enforces Identifiability ​

Author: Mathieu Cyrille Simon, Pascal Frossard, Christophe De Vleeschouwer
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.21385v2 Announce Type: replace Abstract: This paper explores unsupervised disentangled representation learning from a functional perspective. We define latent concepts as factors that influence observations through locally orthogonal directions, formalized as an orthogonality constraint o...

📖 Read original article


433. Decodable but Not Faithful: Coupling Natural-Language Rationales to Programmatic Verifiers ​

Author: Vatsal Ananthula, Adarsh Kumarappan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2606.21678v2 Announce Type: replace Abstract: Language models can generate plausible rationales for their predictions, but these explanations may not faithfully represent the model's internal reasoning. We propose verifier-coupled reasoning, a framework that inserts inline claims into reasonin...

📖 Read original article


434. ATMA: Long-Context Language Modeling via Polar Attention and Gated-Delta Compression Memory ​

Author: Habibullah Akbar
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.25156v3 Announce Type: replace Abstract: Native length extrapolation remain a weakly solvable problem in language modeling due to trade-off balancing between exact retrieval fidelity, long-document likelihood, and inference efficiency. We present ATMA as a disciplined diagnostic for this ...

📖 Read original article


435. Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization ​

Author: Jiading Gai, Shuai Zhang, Kaj Bostrom, Jin Huang, Vihang Patil, Haoyang Fang, Bernie Wang, Huzefa Rangwala, George Karypis
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.26453v2 Announce Type: replace Abstract: We present KernelPro, a closed-loop multi-agent system that automatically generates, profiles, and iteratively optimizes GPU kernel code by integrating large language model (LLM) code generation with hardware profiler feedback and pluggable bottlen...

📖 Read original article


436. DRIFT: Difficulty Routing Self-DIstillation with Rhythm-Gated Exploration and Success BuFfer Training ​

Author: Haisen Luo, Yiwei Liu, Haoning Wang, Dan Liu, Junxi Yin, Haotian Wang, Lei Zhang, Xiaoyu Tian, Shuaiting Chen, Yuansheng Song, Baoyan Guo, Xiongfei Yan, Bolan Yang, Chengwei Liu, Ming Cui, Jiong Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.30345v4 Announce Type: replace Abstract: Enabling large language models to achieve stable self-improvement without external expert supervision remains a central challenge in complex reasoning tasks. Existing self-distillation and reinforcement learning methods lack explicit mechanisms for...

📖 Read original article


437. Visualizing High-Dimensional Graph Embeddings via Informed Multi-View Projections ​

Author: Ya Ji (Khoury College of Computer Sciences, Northeastern University, Seattle), Xuefeng Li (Khoury College of Computer Sciences, Northeastern University, Seattle), Timo Brand (School of Computation, Information and Technology, Technical University of Munich, Heilbronn, Germany), Jacob Miller (School of Computation, Information and Technology, Technical University of Munich, Heilbronn, Germany), Peng Zhang (Khoury College of Computer Sciences, Northeastern University, Seattle), Stephen Kobourov (School of Computation, Information and Technology, Technical University of Munich, Heilbronn, Germany), Yifan Hu (Khoury College of Computer Sciences, Northeastern University, Seattle)
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.31119v2 Announce Type: replace Abstract: Graphs are commonly visualized in 2D, where humans readily interpret spatial relationships, yet such layouts often distort higher-dimensional structure. We propose to embed graphs in high-dimensional space and search for informative 2D viewpoints t...

📖 Read original article


438. NeuroBridge: Bridging Multi-Task MRI Knowledge for Neurodegenerative Disease Diagnosis ​

Author: Mengyu Li, Guoyao Shen, Chad W. Farris, Xin Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.01401v2 Announce Type: replace Abstract: Accurate MRI-based identification of Alzheimer's disease (AD), mild cognitive impairment (MCI), and related dementias remains challenging because disease-related structural changes are often subtle and heterogeneous. We developed NeuroBridge, a cli...

📖 Read original article


439. Full-Stack FP4: Stable LLM Pretraining with Quantized Projections, Optimizers, and Attention ​

Author: Siyu Ding, Mingchuan Ma, Jiabo Tong, Xingrun Xing, Ziming Wang, Guoqi Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.04422v2 Announce Type: replace Abstract: Recent NVFP4 pretraining work has primarily optimized Transformer linear projections, leaving persistent optimizer states, optimizer computation, and low-precision attention forward--backward paths less explored. We present \textbf{Full-Stack FP4},...

📖 Read original article


440. Minimum Block Width for Universal Approximation by Residual Neural Networks with Inner Width One ​

Author: Qi Zhou, Xuan Zhou, Xiao-Song Yang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.04597v3 Announce Type: replace Abstract: In this paper, we study the universal approximation property of residual neural networks. For input and output dimensions $d_x$ and $d_y$, and LeakyReLU, ReLU, ReLU-like activation functions, the upper and lower bounds of the minimum block width ar...

📖 Read original article


441. RSPO: Reward-Swap Policy Optimization for Multi-Turn LLM Agents ​

Author: Qiang Liu, Taian Guo, Ruizhi Qiao, Xing Sun
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.04713v2 Announce Type: replace Abstract: Reinforcement learning holds significant potential for training large language models (LLMs) to handle multi-turn interactive tasks. However, in long-horizon, multi-turn tasks characterized by sparse outcome rewards, directly training with outcome ...

📖 Read original article


442. Safe Bayesian Optimization with Counterfactual Policies ​

Author: Katherine Avery, Bruno Castro da Silva, David Jensen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.05620v2 Announce Type: replace Abstract: In many decision-making settings, new interventions are acceptable only if they do not reduce outcomes below some established threshold. For example, in clinical medicine, new treatments are often acceptable only if they do not worsen outcomes rela...

📖 Read original article


443. Modeling Normal Is All You Need: Joint Latent Clustering for Anomaly Detection in Multimodal Cyber-Physical Systems ​

Author: Alexander Apartsin, Yehudit Aperstein
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.06094v2 Announce Type: replace Abstract: A cyber-physical system (CPS) can enter a faulty state that is individually normal on every sensor and reconstructs accurately, yet is improbable under normal joint operation. This exposes the central weakness of reconstruction-based detection: rec...

📖 Read original article


444. KronQ: LLM Quantization via Kronecker-Factored Hessian ​

Author: Donghyun Lee, Yuhang Li, Ruokai Yin, Priyadarshini Panda
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.07964v2 Announce Type: replace Abstract: Post-training quantization (PTQ) is a widely adopted technique for compressing large language models (LLMs) without retraining. Most existing second-order PTQ methods, including GPTQ, construct quantization objectives from input activation statisti...

📖 Read original article


445. When Counterbalancing Hides the Bias: Access-Conditioned Position Lock in Forced-Choice LLM Evaluation ​

Author: Hong-In Won, Jinseok Jang, Hyoseop Kim
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10202v2 Announce Type: replace Abstract: Forced-choice probes with counterbalanced orientations are a standard tool for measuring language-model "value dispositions," and a concentration/extremity index over repeated draws is read as how sharply a model commits. We show this estimator is ...

📖 Read original article


446. Proxy OPD: On-Policy Distillation with Transferable Relative Proxy Update ​

Author: Daocheng Fu, Rong Wu, Yu Yang, Jianbiao Mei, Licheng Wen, Pinlong Cai, Xuemeng Yang, Yong Liu, Botian Shi, Yu Qiao
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11505v2 Announce Type: replace Abstract: Post-training for large language models typically couples policy exploration with model optimization, hindering the reuse of high-reward behaviors from policy exploration. While on-policy distillation alleviates this by consolidating independently ...

📖 Read original article


447. An Auto-Scaling Approach for Serverless Environments Based on a Multi-Expert Consensus Mechanism ​

Author: Mobina Kashaniyan, Mehrdad Ashtiani, Amirhossein Ghassemi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC, cs.PF

arXiv:2607.15511v2 Announce Type: replace Abstract: Serverless computing provides automatic resource management and pay-per-use execution, but effective autoscaling remains challenging because of dynamic workloads, cold-start latency, and dependencies among functions. We present a dependency-aware a...

📖 Read original article


448. Understanding Reasoning from Pretraining to Post-Training ​

Author: Jingyan Shen, Ang Li, Salman Rahman, Yifan Sun, Micah Goldblum, Matus Telgarsky, Pavel Izmailov
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.16097v2 Announce Type: replace Abstract: Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is largely studied in isolation from the pretraining that precedes it. As a result, two basic questions remain...

📖 Read original article


449. OpenMHC: Accelerating the Science of Wearable Foundation Models ​

Author: Narayan Schuetz, Yuze Bai, Lianggang Pan, Edgar Eggert, Favour Nerrise, Juan Delgado-SanMartin, Max Rosenblattl, Milana Gurbanova, Mohammad Asadi, Anders Johnson, Paul Schmiedmayer, Dennis Wang, Allan Lawrie, Daniel Seung Kim, Xin Liu, Akshay Paruchuri, Ehsan Adeli, Euan Ashley, Kelly W. Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.16235v3 Announce Type: replace Abstract: Mobile and wearable devices offer an unprecedented opportunity for continuous, passive health monitoring and active health coaching. However, the largest wearable datasets are not publicly available for research, and leading wearable foundation mod...

📖 Read original article


450. FlashRT: Agent Harness for Guiding Agents to Deploy Real-Time Multimodal Applications ​

Author: Krish Agarwal, Zhuoming Chen, Yanyuan Qin, Zhenyu Gu, Atri Rudra, Beidi Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.18171v2 Announce Type: replace Abstract: Real-time multimodal applications, including voice agents and interactive video generation, compose heterogeneous models into pipelines whose efficient deployment requires application-specific decisions about placement, streaming, and intra-model p...

📖 Read original article


451. Cost Accounting for Reactive Computational Graphs: Exhaustive Sweeps, Sequential Mutation, and the Backward-Locality Gap ​

Author: Abdallah Khemais (ISITCOM, University of Sousse)
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.18323v2 Announce Type: replace Abstract: Exhaustive site-by-site interventions on a neural network's computational graph -- activation-patching sweeps, circuit-discovery searches, systematic ablation studies -- mutate the graph at every candidate site, and their cost is dominated by recom...

📖 Read original article


452. HindsightBench: A Black-Box Behavioral Audit Protocol for Parametric Hindsight in Time-Indexed LLM Decision Tasks ​

Author: Haozhe Jia
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.18867v2 Announce Type: replace Abstract: Large language models leak parametric knowledge of what followed a historical date into decision tasks indexed by that date -- not necessarily a lookup of the realized outcome, but knowledge of the period all the same. Existence is settled; what us...

📖 Read original article


453. Riemannian Deep Learning: Modules, Networks, and Geometries ​

Author: Ziheng Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DG

arXiv:2607.19305v3 Announce Type: replace Abstract: Deep neural networks on manifold-valued representations have attracted growing interest, but many basic components remain tied to specific manifolds, rely on Euclidean approximations, or require costly and numerically fragile geometric operations. ...

📖 Read original article


454. Multilevel Graph Wavelet Compressed Sensing with Scale-Aware Neural Recovery ​

Author: Amirhossein Nouranizadeh, Sarang Rajendra Patil, Alan John Varghese, Varsha Narayanan, Amit Chakraborty, Mengjia Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20857v2 Announce Type: replace Abstract: Scientific machine learning methods such as neural operators and physics-informed neural networks have advanced engineering applications and inverse problems, but their training typically requires large volumes of simulated data. This makes data pr...

📖 Read original article


455. Dysphagia Risk Stratification in Head and Neck Cancer via Two-Stage PRO-Clinical Stacking ​

Author: Siyuan Zhao, Eric Ababio Anyimadu, Zachary G. Brumm, Yue Ma, Clifton David Fuller, Xinhua Zhang, G. Elisabeta Marai, Guadalupe Canahuate
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.22514v2 Announce Type: replace Abstract: Dysphagia is a debilitating late effect of head and neck cancer (HNC) treatment, yet timely identification of at-risk patients remains challenging in survivorship care. Definitive assessment relies on videofluoroscopic imaging, as captured by the D...

📖 Read original article


456. Wrong Design Intent Is Worse Than Never Conditioning: A Derangement-Control Diagnosis of Header Conditioning in CAD Program Completion ​

Author: Yang Xiao
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.GR

arXiv:2607.23191v3 Announce Type: replace Abstract: Fine-tuned code LLMs are routinely conditioned on a design-intent specification, but the correctness axis of such a signal -- a wrong intent rather than an absent one -- has not been tested, and the benefit of conditioning is usually scored with th...

📖 Read original article


457. From Interface to Inference: Eliciting Any-Order Inference from Any-Order Models ​

Author: Seunggeun Kim, Jaeyeon Kim, Taekyun Lee, Yuyuan Chen, Yilun Du, Sham Kakade, Sitan Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.26504v2 Announce Type: replace Abstract: Many discrete reasoning tasks, such as code generation, are inherently non-causal: programmers move between high-level structure and local details, a process we call any-order inference. For autoregressive language models, which lack a native any-o...

📖 Read original article


458. Hierarchical Copula-Gumbel-Top-\texorpdfstring{$K$}{K} Routing: Two-Sided Dependence Control for Frozen Mixture-of-Experts at Fixed Per-Token Routing Laws ​

Author: Richard Yi Da Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.28670v2 Announce Type: replace Abstract: A stochastic Gumbel-Top-$K$ router defines, for every token of a mixture-of-experts (MoE) model, a \emph{routing law}: a distribution over ordered expert lists and mixture weights. We ask which \emph{joint} distributions over the routing choices of...

📖 Read original article


459. End-to-End Fairness Optimization with Fair Decision-Focused Learning ​

Author: Yu Wang, Violet Xinying Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2607.29441v2 Announce Type: replace Abstract: Many real-world systems rely on predictive models to inform decisions, and fairness concerns arise in both the prediction and decision stages. We introduce end-to-end fairness optimization (E2EFO) as a unifying framework that integrates fairness ac...

📖 Read original article


460. Factorized AdaBoost.MH Achieves the Same Convergence Rate as AdaBoost.MH ​

Author: Xin Zou, Jingyuan Xu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.01091v2 Announce Type: replace Abstract: {AdaBoost.MH} reduces multi-class classification to a collection of binary subproblems and enjoys the classical boosting-type convergence guarantee under a weak learning condition. A more structured variant, Factorized {AdaBoost.MH}, uses base clas...

📖 Read original article


461. On the Identifiability of Masked Prediction: Mode Blindness and Mask Schedules ​

Author: Yichao Cai, Javen Qinfeng Shi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT

arXiv:2608.01383v2 Announce Type: replace Abstract: Masked prediction learns representations by fitting a schedule-weighted family of conditional laws, but it remains unclear when near-optimal conditional prediction pins down the underlying joint law. We study this question for data with two well-se...

📖 Read original article


462. RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States ​

Author: Yi Yang, Zhennan Chen, Yihong Zhuang, Tiehan Fan, Yinan Chen, Jian Li, Jian Yang, Ying Tai
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.02508v3 Announce Type: replace Abstract: Learning-based memory systems for self-evolving LLM agents face two tightly coupled challenges. First, trajectory-indexed utilities grow with the interaction history, thereby dispersing limited feedback over an ever-expanding state space. Second, b...

📖 Read original article


463. GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning ​

Author: Zhaoxin Yu, Qi Shen, Hengli Li, Zhaowei Zhang, Song-Chun Zhu, Chi Zhang, Zilong Zheng
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.02585v2 Announce Type: replace Abstract: Optimization-based latent reasoning improves large language model outputs by optimizing instance-specific continuous states at test time while keeping model parameters frozen. Existing methods, however, typically connect these states to the reasoni...

📖 Read original article


464. Resume Means Resume: A Machine-Checked Conformance Contract for Checkpoint, Interrupt, and Resume Semantics in Workflow Persistence Layers ​

Author: Sajjad Khan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.LO, cs.SE

arXiv:2608.03836v3 Announce Type: replace Abstract: A framework that persists execution state so a run can be interrupted, survive a crash, and continue must decide what a resume means for effects that already happened. Five widely deployed agent workflow frameworks answer differently, none exposes ...

📖 Read original article


465. A Trust-region Framework for Moment Estimation ​

Author: Oluwasegun A. Somefun
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SP, eess.SY

arXiv:2608.04026v2 Announce Type: replace Abstract: In this paper, we develop a trust-region framework for understanding the behavior of adaptive moment estimation mechanisms, such as \textsc{Adam}, in stochastic gradient optimization. Specifically, the magnitude of the update step associated with e...

📖 Read original article


466. ArborEnum: Decision Tree Rashomon Sets over Continuous Features ​

Author: Zakk Heile, Hayden McTavish, Margo Seltzer, Cynthia Rudin
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.04310v2 Announce Type: replace Abstract: The Rashomon effect describes the phenomenon that many models can achieve nearly equivalent performance on the same learning task, with significant ramifications for robustness, feature importance, and customizability. These use cases motivate the ...

📖 Read original article


467. On-Policy Self-Distillation without Any Supervision ​

Author: Yijiang Li, Bingyang Wang, Yijun Liang, Yunjie Tian, Di Fu, Nuno Vasconcelos
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.06296v2 Announce Type: replace Abstract: On-policy (Self-)Distillation (OPD / OPSD) has shown strong potential for post-training large language models (LLMs). However, existing methods still rely heavily on external supervision, including ground-truth signals, environmental feedback, or g...

📖 Read original article


468. Latent Fact-Checking: Detecting Misinformation through Activation Engineering ​

Author: Pedro Barcelos, Ot'avio Parraga, Marcelo M. Mussi, Lucas M. Fraga, Lucas S. Kupssinsk"u, Rodrigo C. Barros
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.06417v2 Announce Type: replace Abstract: The proliferation of misinformation online has driven demand for scalable detection systems. While most existing approaches rely on surface-level linguistic features or external knowledge retrieval, we examine truthfulness as a geometric property o...

📖 Read original article


469. Which Decisions Low-Bit Quantization Breaks, and How to Predict Them ​

Author: Zekun Wu, Swati Dhiman, Adriano Koshiyama
Published: 8/11/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.06564v2 Announce Type: replace Abstract: Quantization is how large language models are actually deployed, and below four bits it hurts. What nobody can say is which decisions change at a given bit-width -- which matters most where a model acts rather than answers, since a tool call it dec...

📖 Read original article


470. The Role of Pseudo-labels in Self-training Linear Classifiers on High-dimensional Gaussian Mixture Data ​

Author: Takashi Takahashi
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.dis-nn, cond-mat.stat-mech, cs.LG, math.ST, stat.TH

arXiv:2205.07739v5 Announce Type: replace-cross Abstract: Self-training (ST) is a simple yet effective semi-supervised learning method. However, why and how ST improves generalization performance by using potentially erroneous pseudo-labels is still not well understood. To deepen the understanding o...

📖 Read original article


471. Causal Falsification of Digital Twins ​

Author: Rob Cornish, Muhammad Faaiz Taufiq, Arnaud Doucet, Chris Holmes
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ME, cs.CE, cs.LG, stat.AP

arXiv:2301.07210v5 Announce Type: replace-cross Abstract: Digital twins are simulation-based models designed to predict how a real-world process will evolve in response to interventions. This modelling paradigm holds substantial promise in many applications, but rigorous procedures for assessing the...

📖 Read original article


472. Stochastic Subgradient Methods with Guaranteed Global Stability in Nonsmooth Nonconvex Optimization ​

Author: Nachuan Xiao, Xiaoyin Hu, Kim-Chuan Toh
Published: 8/11/2026, 4:00:00 AM
Categories: math.OC, cs.AI, cs.LG, stat.ML

arXiv:2307.10053v5 Announce Type: replace-cross Abstract: In this paper, we focus on providing convergence guarantees for stochastic subgradient methods in minimizing nonsmooth nonconvex functions. We first investigate the global stability of a general framework for stochastic subgradient methods, w...

📖 Read original article


473. A Differentially Private Weighted Empirical Risk Minimization Procedure and its Application to Outcome Weighted Learning ​

Author: Spencer Giddens, Yiwang Zhou, Kevin R. Krull, Tara M. Brinkman, Peter X. K. Song, Fang Liu
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2307.13127v4 Announce Type: replace-cross Abstract: Data used to train predictive models via empirical risk minimization (ERM) often contain sensitive personal information. While differential privacy (DP) provides mathematically provable bounds to protect such data, previous work has focused a...

📖 Read original article


474. From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition ​

Author: Maan Qraitem, Kate Saenko, Bryan A. Plummer
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2308.04553v4 Announce Type: replace-cross Abstract: Visual recognition models are prone to learning spurious correlations induced by a biased training set where certain conditions $B$ (\eg, Indoors) are over-represented in certain classes $Y$ (\eg, Big Dogs). Synthetic data from off-the-shelf ...

📖 Read original article


475. Matrix Completion via Nonsmooth Regularization of Fully Connected Neural Networks ​

Author: Sajad Faramarzi, Farzan Haddadi, Sajjad Amini, Masoud Ahookhosh, Symeon Chatzinotas
Published: 8/11/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT

arXiv:2403.10232v2 Announce Type: replace-cross Abstract: Conventional matrix completion methods approximate the missing values by assuming the matrix to be low-rank, which leads to a linear approximation of missing values. It has been shown that enhanced performance could be attained by using nonli...

📖 Read original article


476. Unsupervised Point Cloud Registration with Self-Distillation ​

Author: Christian L"owens, Thorben Funke, Andr'e Wagner, Alexandru Paul Condurache
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO

arXiv:2409.07558v2 Announce Type: replace-cross Abstract: Rigid point cloud registration is a fundamental problem and highly relevant in robotics and autonomous driving. Nowadays deep learning methods can be trained to match a pair of point clouds, given the transformation between them. However, thi...

📖 Read original article


477. Machine Learning for Inverse Problems and Data Assimilation ​

Author: Eviatar Bach, Ricardo Baptista, Daniel Sanz-Alonso, Andrew Stuart
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC

arXiv:2410.10523v3 Announce Type: replace-cross Abstract: The aim of this book is to demonstrate the potential for ideas in machine learning to impact on the fields of inverse problems and data assimilation. The perspective is one that is primarily aimed at researchers from inverse problems and/or d...

📖 Read original article


478. Error Bound Analysis for the Regularized Loss of Deep Linear Neural Networks ​

Author: Po Chen, Rujun Jiang, Peng Wang
Published: 8/11/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2502.11152v4 Announce Type: replace-cross Abstract: The optimization foundations of deep linear networks have recently received significant attention. However, due to their inherent non-convexity and hierarchical structure, analyzing the loss functions of deep linear networks remains a challen...

📖 Read original article


479. Artifact detection and localization in single-channel mobile EEG for sleep research using deep learning and attention mechanisms ​

Author: Khrystyna Semkiv, Jia Zhang, Maria Laura Ferster, Walter Karlen
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2504.08469v4 Announce Type: replace-cross Abstract: Current methods for detecting artifacts in sleep EEG range from threshold-based algorithms to machine learning approaches, yet applications remain limited for single-channel mobile EEG. We propose a convolutional neural network (CNN) model in...

📖 Read original article


480. Wasserstein Distributionally Robust Regret Optimization ​

Author: Lukas-Benedikt Fiechtner, Jose Blanchet
Published: 8/11/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2504.10796v5 Announce Type: replace-cross Abstract: Distributionally robust optimization (DRO) is widely used for decision-making under uncertainty, but its adversarial focus on worst-case loss can lead to overly conservative policies. To mitigate this, we study ex-ante Distributionally Robust...

📖 Read original article


481. TreeHop: Efficient Embedding-Level Query Rewriter ​

Author: Zhonghao Li, Kunpeng Zhang, Jinghuai Ou, Shuliang Liu, Xuming Hu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.HC, cs.LG

arXiv:2504.20114v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems face significant challenges in multi-hop question answering (MHQA), where complex queries require synthesizing information across multiple document chunks. Existing approaches typically rely on ite...

📖 Read original article


482. On the expressivity of deep Heaviside networks ​

Author: Insung Kong, Juntong Chen, Sophie Langer, Johannes Schmidt-Hieber
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.NA

arXiv:2505.00110v2 Announce Type: replace-cross Abstract: We show that deep Heaviside networks (DHNs) have limited expressiveness but that this can be overcome by including either skip connections or neurons with linear activation. We provide lower and upper bounds for the Vapnik-Chervonenkis (VC) d...

📖 Read original article


483. neuralGAM: An R Package for Fitting Generalized Additive Neural Networks ​

Author: Ines Ortega-Fernandez, Marta Sestelo
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO, stat.ME

arXiv:2505.08610v2 Announce Type: replace-cross Abstract: Nowadays, Neural Networks are considered one of the most effective methods for various tasks such as anomaly detection, computer-aided disease detection, or natural language processing. However, these networks suffer from the ``black-box'' pr...

📖 Read original article


484. Temporal Convolutional Autoencoder for Interference Mitigation in FMCW Radar Altimeters ​

Author: Charles E. Thornton, Jamie Sloop, Samuel Brown, Aaron Orndorff, William C. Headley, Stephen Young
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2505.22783v3 Announce Type: replace-cross Abstract: Reliable altitude estimation with frequency-modulated continuous wave (FMCW) radar altimeters is increasingly a challenge due to in-band interference from modern communication systems. In this paper, we present a temporal convolutional autoen...

📖 Read original article


485. EgoBrain: Synergizing Minds and Eyes For Human Action Understanding ​

Author: Nie Lin, Yansen Wang, Dongqi Han, Weibang Jiang, Jingyuan Li, Ryosuke Furuta, Yoichi Sato, Dongsheng Li
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG

arXiv:2506.01353v3 Announce Type: replace-cross Abstract: The integration of brain-computer interfaces (BCIs), in particular electroencephalography (EEG), with artificial intelligence (AI) has shown tremendous promise in decoding human cognition and behavior from neural signals. In particular, the r...

📖 Read original article


486. WebChoreArena: Evaluating Web Browsing Agents on Realistic Tedious Web Tasks ​

Author: Atsuyuki Miyai, Zaiying Zhao, Kazuki Egashira, Atsuki Sato, Tatsumi Sunada, Shota Onohara, Hiromasa Yamanishi, Mashiro Toyooka, Kunato Nishina, Ryoma Maeda, Kiyoharu Aizawa, Toshihiko Yamasaki
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2506.01952v2 Announce Type: replace-cross Abstract: Powered by large language models (LLMs), web browsing agents operate graphical user interfaces in a human-like manner, offering a transparent and general framework for automating web-based tasks. As these agents rapidly improve and achieve st...

📖 Read original article


487. A Control Function Framework for Mitigating Position Bias in Learning to Rank Systems ​

Author: Md Aminul Islam, Kathryn Vasilaky, Elena Zheleva
Published: 8/11/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2506.06989v3 Announce Type: replace-cross Abstract: Learning-to-rank (LTR) systems commonly depend on implicit feedback, such as user clicks, because it is easy to collect and can serve as a valuable signal of user preferences. However, directly optimizing ranking models using implicit feedbac...

📖 Read original article


488. EEG Foundation Challenge: From Cross-Task to Cross-Subject EEG Decoding ​

Author: Bruno Aristimunha, Dung Truong, Pierre Guetschel, Seyed Yahya Shirazi, Isabelle Guyon, Alexandre R. Franco, Michael P. Milham, Aviv Dotan, Scott Makeig, Alexandre Gramfort, Jean-Remi King, Marie-Constance Corsi, Pedro A. Vald'es-Sosa, Amit Majumdar, Alan Evans, Terrence J Sejnowski, Oren Shriki, Sylvain Chevallier, Arnaud Delorme
Published: 8/11/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2506.19141v3 Announce Type: replace-cross Abstract: Current electroencephalogram (EEG) decoding models are typically trained on small numbers of subjects performing a single task. Here, we introduce a large-scale, code-submission-based competition comprising two challenges. First, the Transfer...

📖 Read original article


489. High-Layer Attention Pruning with Rescaling ​

Author: Songtao Liu, Peng Liu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2507.01900v3 Announce Type: replace-cross Abstract: Pruning is a highly effective approach for compressing large language models (LLMs), significantly reducing inference latency. However, conventional training-free structured pruning methods often employ a heuristic metric that indiscriminatel...

📖 Read original article


490. NeuralDMD: Interpretable Neural Representation of Dynamics from Sparse and Noisy Measurements ​

Author: Ali SaraerToosi, Renbo Tu, Esther Y. H. Lin, Kamyar Azizzadenesheli, Aviad Levis
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, astro-ph.IM, cs.LG, physics.ao-ph

arXiv:2507.03094v2 Announce Type: replace-cross Abstract: Many challenges in scientific imaging involve solving ill-posed inverse problems, where the goal is to recover spatio-temporal fields from indirect, noisy, and highly sparse measurements - often without access to ground truth data or reliable...

📖 Read original article


491. MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks ​

Author: Adrian Marius Dumitran, Theodor-Pierre Moroianu, Mihnea-Vicentiu Buca
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CL, cs.LG

arXiv:2507.03162v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has transformed various domains, particularly computer science (CS) education. These models exhibit remarkable capabilities in code-related tasks and problem-solving, raising questions abo...

📖 Read original article


492. MetaLint: Easy-to-Hard Generalization for Code Linting ​

Author: Atharva Naik, Lawanya Baghel, Dhakshin Govindarajan, Darsh Agrawal, Yiqing Xie, Daniel Fried, Carolyn Rose
Published: 8/11/2026, 4:00:00 AM
Categories: cs.SE, cs.CL, cs.LG

arXiv:2507.11687v5 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practices beyond those observed during training. We introduce MetaLint, a meta-learning framework that form...

📖 Read original article


493. LogicIF: Towards Complex Logic Instruction Following ​

Author: Mian Zhang, Shujian Liu, Sixun Dong, Ming Yin, Yebowen Hu, Xun Wang, Simin Ma, Song Wang, Sathish Reddy Indurthi, Haoyun Deng, Zhiyu Zoey Chen, Kaiqiang Song
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2508.09125v4 Announce Type: replace-cross Abstract: Instruction following has catalyzed the recent era of Large Language Models (LLMs) and is the foundational skill underpinning more advanced capabilities such as reasoning and agentic behaviors. As tasks grow more challenging, the logic struct...

📖 Read original article


494. Efficient identification of critical regions via Flow Matching-based Monte Carlo initialization ​

Author: Qian-Rui Lee, Daw-Wei Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cond-mat.stat-mech, cs.LG

arXiv:2508.15318v5 Announce Type: replace-cross Abstract: Markov chain Monte Carlo (MCMC) is a standard tool for studying many-body systems, but its practical cost can become substantial, especially when simulations must be repeated across temperatures and lattice sizes or near transition regions wh...

📖 Read original article


495. ML-PWS: Estimating the Mutual Information Between Experimental Time Series Using Neural Networks ​

Author: Manuel Reinhardt, Ga\v{s}per Tka\v{c}ik, Pieter Rein ten Wolde
Published: 8/11/2026, 4:00:00 AM
Categories: physics.bio-ph, cond-mat.stat-mech, cs.IT, cs.LG, math.IT, q-bio.NC

arXiv:2508.16509v3 Announce Type: replace-cross Abstract: The ability to quantify information transmission is crucial for the analysis and design of both natural and engineered systems. For systems driven by time-varying signals, the fundamental measure is the information transmission rate. However,...

📖 Read original article


496. Enhancing Knowledge Tracing through Leakage-Free and Recency-Aware Embeddings ​

Author: Yahya Badran, Christine Preisach
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG

arXiv:2508.17092v2 Announce Type: replace-cross Abstract: Knowledge Tracing (KT) aims to predict a student's future performance based on their sequence of interactions with learning content. Many KT models rely on knowledge concepts (KCs), which represent the skills required for each item. However, ...

📖 Read original article


497. An invertible generative model for forward and inverse problems ​

Author: Christoph Brune, Marcello Carioni, Tristan van Leeuwen, Lasse Veenstra
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR

arXiv:2509.03910v2 Announce Type: replace-cross Abstract: We formulate inverse problems in a Bayesian framework and aim to train an invertible generative model that is capable of simulation (i.e., sampling from the likelihood) and inference (i.e., sampling from the posterior). We call such a generat...

📖 Read original article


498. Scalable extensions to given-data Sobol' index estimators ​

Author: Teresa Portone, Bert Debusschere, Samantha Yang, Emiliano Islas-Quinones, T. Patrick Xiao
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP, stat.CO

arXiv:2509.09078v3 Announce Type: replace-cross Abstract: Given-data methods for variance-based sensitivity analysis have significantly advanced the feasibility of Sobol' index computation for computationally expensive models and models with many inputs. However, the limitations of existing methods ...

📖 Read original article


499. Matrix-free Neural Preconditioner for the Dirac Operator in Lattice Gauge Theory ​

Author: Yixuan Sun, Srinivas Eswar, Yin Lin, William Detmold, Phiala Shanahan, Xiaoye Li, Yang Liu, Prasanna Balaprakash
Published: 8/11/2026, 4:00:00 AM
Categories: hep-lat, cs.LG

arXiv:2509.10378v2 Announce Type: replace-cross Abstract: Linear systems arise in generating samples and in calculating observables in lattice quantum chromodynamics~(QCD). Solving the Hermitian positive definite systems, which are sparse but ill-conditioned, involves using iterative methods, such a...

📖 Read original article


500. Quokka: Accelerating Program Verification with LLMs via Invariant Synthesis ​

Author: Anjiang Wei, Tianran Sun, Tarun Suresh, Haoze Wu, Ke Wang, Alex Aiken
Published: 8/11/2026, 4:00:00 AM
Categories: cs.PL, cs.AI, cs.CL, cs.LG

arXiv:2509.21629v4 Announce Type: replace-cross Abstract: Program verification relies on loop invariants, yet automatically discovering strong invariants remains a long-standing challenge. We investigate whether large language models (LLMs) can accelerate program verification by generating useful lo...

📖 Read original article


501. ToolUniverse: An open platform for democratizing AI scientists ​

Author: Shanghua Gao, Richard Zhu, Pengwei Sui, Zhenglun Kong, Sufian Aldogom, Yepeng Huang, Ayush Noori, Reza Shamji, Krishna Parvataneni, Theodoros Tsiligkaridis, Marinka Zitnik
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2509.23426v3 Announce Type: replace-cross Abstract: AI scientists are emerging computational systems that serve as collaborative partners in discovery. These systems remain difficult to build because they are bespoke, tied to rigid workflows, and lack shared environments that unify tools, data...

📖 Read original article


502. MoE-Prism: Disentangling Monolithic Experts for Elastic MoE Services via Model-System Co-Designs ​

Author: Xinfeng Xia, Xiaofeng Hou, Jiacheng Liu, Wenfeng Wang, Mingxuan Zhang, Peng Tang, Chao Li, Minyi Guo
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2510.19366v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) scales model capacity through sparse activation, and is becoming an important architecture for large language models (LLMs). However, existing MoE serving systems typically execute all requests under a fixed routing c...

📖 Read original article


503. Learning to Triage Vulnerability Reports from Program Analysis: An Empirical Study in Node.js ​

Author: Ronghao Ni, Aidan Z. H. Yang, Min-Chien Hsu, Nuno Sabino, Limin Jia, Ruben Martins, Darion Cassel, Kevin Cheang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.SE

arXiv:2510.20739v2 Announce Type: replace-cross Abstract: Program analysis tools often produce large volumes of candidate vulnerability reports that require costly manual review, creating a practical challenge: how can security analysts prioritize the reports most likely to be true vulnerabilities? ...

📖 Read original article


504. Embedding Trust: Semantic Isotropy Predicts Nonfactuality in Long-Form Text Generation ​

Author: Dhrupad Bhardwaj, Julia Kempe, Tim G. J. Rudner
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, stat.ME, stat.ML

arXiv:2510.21891v2 Announce Type: replace-cross Abstract: To deploy large language models (LLMs) in high-stakes application domains that require substantively accurate responses to open-ended prompts, we need reliable, computationally inexpensive methods that assess the trustworthiness of long-form ...

📖 Read original article


505. QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture ​

Author: Shvetank Prakash, Andrew Cheng, Mark Mazumder, Arya Tschand, Varun Gohil, Jeffrey Ma, Jason Yik, Zishen Wan, Jessica Quaye, Elisavet Lydia Alvanaki, Avinash Kumar, Chandrashis Mazumdar, Tuhin Khare, Alexander Ingare, Ikechukwu Uchendu, Radhika Ghosal, Abhishek Tyagi, Chenyu Wang, Andrea Mattia Garavagno, Sarah Gu, Alice Guo, Grace Hur, Luca P. Carloni, Tushar Krishna, Ankita Nayak, Amir Yazdanbakhsh, Vijay Janapa Reddi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AR, cs.AI, cs.LG, cs.SE

arXiv:2510.22087v3 Announce Type: replace-cross Abstract: The field of computer architecture, which bridges high-level software abstractions and low-level hardware implementations, remains absent from current large language model (LLM) evaluations. To this end, we present QuArch (pronounced 'quark')...

📖 Read original article


506. Floating-Point Neural Network Verification at the Software Level ​

Author: Edoardo Manino, Bruno Farias, Rafael S'a Menezes, Fedor Shmarov, Lucas C. Cordeiro
Published: 8/11/2026, 4:00:00 AM
Categories: cs.SE, cs.CR, cs.LG

arXiv:2510.23389v2 Announce Type: replace-cross Abstract: The behaviour of neural network components must be proven correct before deployment in safety-critical systems. Unfortunately, existing neural network verification techniques cannot certify the absence of faults at the software level. In this...

📖 Read original article


507. Assessing Factual Music Comprehension in Large Audio Language Models ​

Author: Daniel Chenyu Lin, Michael Freeman, John Thickstun
Published: 8/11/2026, 4:00:00 AM
Categories: cs.SD, cs.CL, cs.LG

arXiv:2511.05550v3 Announce Type: replace-cross Abstract: Large audio language models (LALMs) leverage multimodal representations to generate open-ended answers to natural language queries about audio. In this paper, we (1) provide empirical evidence that assessment of LALMs using the popular MusicQ...

📖 Read original article


Author: Adam \v{S}torek, Vikas Upadhyay, Marianne Menglin Liu, Daniel W. Peterson, Anshul Mittal, Sujeeth Bharadwaj, Fahad Shah, Sujith Ravi, Dan Roth
Published: 8/11/2026, 4:00:00 AM
Categories: cs.SE, cs.CL, cs.LG

arXiv:2511.09373v2 Announce Type: replace-cross Abstract: LLMs now tackle a wide range of software-related tasks, yet we show that their performance varies markedly both across and within these tasks. Routing user queries to the appropriate LLMs can therefore help improve response quality while redu...

📖 Read original article


509. Sparse corruption in low-rank matrix inference: the PCA benchmark ​

Author: Urte Adomaityte, Gabriele Sicuro, Pierpaolo Vivo
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.dis-nn, cond-mat.stat-mech, cs.LG

arXiv:2511.11927v2 Announce Type: replace-cross Abstract: Principal Component Analysis (PCA) is a standard tool for extracting a low-rank signal from noisy observations. It is known that applying PCA to a rank-one signal corrupted by a dense, homogeneous noise, in the large matrix size limit, the ce...

📖 Read original article


510. Tokenisation over Bounded Alphabets is Hard ​

Author: Violeta Kastreva, Philip Whittington, Dennis Komm, Tiago Pimentel
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.DS, cs.LG

arXiv:2511.15709v2 Announce Type: replace-cross Abstract: Recent works have shown that tokenisation is NP-complete. However, these works assume tokenisation is applied to inputs with unboundedly large alphabets -- an unrealistic assumption, given that in practice tokenisers operate over fixed-size a...

📖 Read original article


511. Integrating RCTs, RWD, AI/ML and Statistics: Next-Generation Evidence Synthesis ​

Author: Shu Yang, Margaret Gamalo, Haoda Fu
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ME, cs.LG

arXiv:2511.19735v2 Announce Type: replace-cross Abstract: Randomized controlled trials (RCTs)have been the cornerstone of clinical evidence; however, their cost, duration, and restrictive eligibility criteria limit power and external validity. Studies using real-world data (RWD), historically consid...

📖 Read original article


512. Length-MAX Tokenizer for Language Models ​

Author: Dong Dong, Weijie Su
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2511.20849v2 Announce Type: replace-cross Abstract: We introduce a new tokenizer for language models that minimizes the average tokens per character, thereby reducing the number of tokens needed to represent text during training and to generate text during inference. Our method, which we refer...

📖 Read original article


513. SIP: Site in Pieces- A Dataset of Disaggregated Construction-Phase 3D Scans for Semantic Segmentation and Scene Understanding ​

Author: Seongyong Kim, Yong Kwon Cho
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2512.09062v2 Announce Type: replace-cross Abstract: Accurate 3D scene interpretation in active construction sites is essential for progress monitoring, safety assessment, and digital twin development. LiDAR is widely used in construction because it offers advantages over camera-based systems, ...

📖 Read original article


514. Neuronal Attention Circuit (NAC) for Representation Learning ​

Author: Waleed Razzaq, Yun-Bo Zhao
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2512.10282v4 Announce Type: replace-cross Abstract: Attention improves representation learning over RNNs, but its discrete nature limits continuous-time (CT) modeling. We introduce Neuronal Attention Circuit (NAC), a novel, biologically inspired CT-attention mechanism that reformulates attenti...

📖 Read original article


515. Bounding Hallucinations: Merlin-Arthur Protocols for Mutual-Information Bounds in Language Models ​

Author: Bj"orn Deiseroth, Max Henning H"oth, Kristian Kersting, Letitia Parcalabescu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2512.11614v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) relies on retrieved context to guide large language models (LLM), yet treats the retrieval as a heuristic rather than verifiable evidence -- leading to unsupported answers, hallucinations, and reliance on ...

📖 Read original article


516. General OOD Detection via Model-aware and Subspace-aware Variable Priority ​

Author: Min Lu, Hemant Ishwaran
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2512.13003v2 Announce Type: replace-cross Abstract: Out-of-distribution (OOD) detection is essential for determining when a supervised model encounters inputs that differ meaningfully from its training distribution. While widely studied in classification, OOD detection for regression and survi...

📖 Read original article


517. FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback ​

Author: Xueqing Wu, Zihan Xue, Da Yin, Shuyan Zhou, Kai-Wei Chang, Nanyun Peng, Yeming Wen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.CV, cs.LG, cs.SE

arXiv:2601.04203v3 Announce Type: replace-cross Abstract: We present FronTalk, a benchmark for front-end code generation that pioneers the study of a unique interaction dynamic: conversational code generation with multi-modal feedback. In front-end development, visual artifacts such as sketches, moc...

📖 Read original article


518. Communication-efficient distributed hazard difference estimation for heterogeneous multi-site survival data ​

Author: Ziwen Wang, Siqi Li, Marcus Eng Hock Ong, Nan Liu
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2601.14609v2 Announce Type: replace-cross Abstract: Multi-site collaboration can power survival models that no single hospital could fit alone, but privacy rules and protected computing environments block patient-level data sharing and the persistent server connections required by iterative fe...

📖 Read original article


519. Demystifying Prediction Powered Inference ​

Author: Yilin Song, Dan M. Kluger, Harsh Parikh, Tian Gu
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2601.20819v2 Announce Type: replace-cross Abstract: Machine learning predictions are increasingly used to supplement incomplete or costly-to-measure outcomes in fields such as biomedical research, environmental science, and social science. However, treating predictions as ground truth introduc...

📖 Read original article


520. Position: Certifiable State Integrity Should Be Built from Local Validity, Not Global Scale ​

Author: Enzo Nicol'as Spotorno, Joao R. Campos, Ant^onio Augusto Medeiros Fr"ohlich
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2601.21249v2 Announce Type: replace-cross Abstract: Breakthroughs in language and vision have motivated increasingly general foundation models for time series and physical dynamics, where evidence is promising but less mature. In safety-critical Cyber-Physical Systems (CPS), globally parameter...

📖 Read original article


521. A solvable high-dimensional model where nonlinear autoencoders learn structure invisible to PCA while test loss misaligns with generalization ​

Author: Vicente Conde Mendes, Lorenzo Bardone, C'edric Koller, Jorge Medina Moreira, Vittorio Erba, Emanuele Troiani, Lenka Zdeborov'a
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.dis-nn, cs.LG

arXiv:2602.10680v2 Announce Type: replace-cross Abstract: Many real-world datasets contain hidden structure that cannot be detected by simple linear correlations between input features. For example, latent factors may influence the data in a coordinated way, even though their effect is invisible to ...

📖 Read original article


522. HyperDet: 3D Object Detection with Hyper 4D Radar Point Clouds ​

Author: Yichun Xiao, Runwei Guan, Jin Jin, Fangqiang Ding
Published: 8/11/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2602.11554v4 Announce Type: replace-cross Abstract: How far can 3D object detection go using 4D radar alone? Despite offering weather-robust and velocity-aware sensing for autonomous perception, modern 4D radar still yields sparse, noisy, and unstable point clouds, limiting radar-only 3D detec...

📖 Read original article


523. SPD Learn: A Geometric Deep Learning Python Library for Neural Decoding Through Trivialization ​

Author: Bruno Aristimunha, Ce Ju, Antoine Collas, Florent Bouchard, Ammar Mian, Bertrand Thirion, Sylvain Chevallier, Reinmar Kobler
Published: 8/11/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG

arXiv:2602.22895v2 Announce Type: replace-cross Abstract: Implementations of symmetric positive definite (SPD) matrix-based neural networks for neural decoding remain fragmented across research codebases and Python packages. Existing implementations often employ ad hoc handling of manifold constrain...

📖 Read original article


524. Variance reduction in lattice QCD observables via normalizing flows ​

Author: Ryan Abbott, Denis Boyda, Yang Fu, Daniel C. Hackett, Gurtej Kanwar, Fernando Romero-L'opez, Phiala E. Shanahan, Julian M. Urban
Published: 8/11/2026, 4:00:00 AM
Categories: hep-lat, cs.LG

arXiv:2603.02984v2 Announce Type: replace-cross Abstract: Normalizing flows can be used to construct unbiased, reduced-variance estimators for lattice field theory observables that are defined by a derivative with respect to action parameters. This work implements the approach for observables involv...

📖 Read original article


525. Design Space of Self--Consistent Electrostatic Machine Learning Interatomic Potentials ​

Author: William J. Baldwin, Ilyes Batatia, Martin Vondr'ak, Johannes T. Margraf, G'abor Cs'anyi
Published: 8/11/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG

arXiv:2603.14700v2 Announce Type: replace-cross Abstract: Machine learning interatomic potentials (MLIPs) have become widely used tools in atomistic simulations. For much of the history of this field, the most commonly employed architectures were based on short-ranged atomic energy contributions, an...

📖 Read original article


526. SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs ​

Author: Yadi Cao, Sicheng Lai, Jiahe Huang, Yang Zhang, Zach Lawrence, Rohan Bhakta, Izzy F. Thomas, Mingyun Cao, Chung-Hao Tsai, Zihao Zhou, Yidong Zhao, Hao Liu, Alessandro Marinoni, Alexey Arefiev, Rose Yu
Published: 8/11/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.AI, cs.DC, cs.LG

arXiv:2603.20253v3 Announce Type: replace-cross Abstract: Evaluating LLM agents for scientific tasks has focused on token costs while ignoring tool-use costs like simulation time and experimental resources. As a result, metrics like pass@k become impractical under realistic budget constraints. To ad...

📖 Read original article


527. CGRL: Causal-Guided Representation Learning for Node-Level Out-of-Distribution Generalization ​

Author: Bowen Lu, Liangqiang Yang, Teng Li, Kun Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2603.24304v2 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) deliver strong performance on graph tasks, but their accuracy drops significantly under out-of-distribution (OOD) scenarios. Under distribution shifts, GNNs often fit environmental noise and spurious correlations ...

📖 Read original article


528. A Comparative Study in Surgical AI: Potential and Limitations of Data, Compute, and Scaling ​

Author: Kirill Skobelev, Eric Fithian, Yegor Baranovski, Jack Cook, Sandeep Angara, Shauna Otto, Zhuang-Fang Yi, John Zhu, Neeraj Mainkar, Margaux Masson-Forsythe, Daniel A. Donoho, X. Y. Han
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG

arXiv:2603.27341v4 Announce Type: replace-cross Abstract: Recent Artificial Intelligence (AI) models have matched or exceeded human experts in several benchmarks of biomedical task performance, but surgical benchmarks in particular are often missing from prominent medical benchmark suites. Since sur...

📖 Read original article


529. To Memorize or to Retrieve: Scaling the Interaction Between Pretraining and Retrieval ​

Author: Karan Singh, Michael Yu, Varun Gangal, Zhuofu Tao, Sachin Kumar, Emmy Liu, Steven Y. Feng
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.00715v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) improves language model (LM) performance by providing relevant context at test time for knowledge-intensive situations. In this work, we systematically study the trade-off between pretraining and retrieval...

📖 Read original article


530. Characterization of Gaussian Universality Breakdown in High-Dimensional Empirical Risk Minimization ​

Author: Chiheb Yaakoubi, Cosme Louart, Malik Tiomoko, Zhenyu Liao
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2604.03146v4 Announce Type: replace-cross Abstract: We study high-dimensional convex empirical risk minimization (ERM) under general non-Gaussian data designs. By heuristically extending the Convex Gaussian Min-Max Theorem (CGMT) to non-Gaussian settings, we derive an asymptotic min-max charac...

📖 Read original article


531. TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories ​

Author: Yen-Shan Chen, Sian-Yao Huang, Cheng-Lin Yang, Yun-Nung Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG, cs.SE

arXiv:2604.07223v2 Announce Type: replace-cross Abstract: As large language models (LLMs) evolve from static chatbots into autonomous agents, the primary vulnerability surface shifts from final outputs to intermediate execution traces. While safety guardrails are well-benchmarked for natural languag...

📖 Read original article


532. Compatibility of Face Embeddings Across Deep Neural Networks ​

Author: Fizza Rubab, Yiying Tong, Arun Ross
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2604.07282v2 Announce Type: replace-cross Abstract: Automated face recognition has made rapid strides over the past decade due to the unprecedented rise of deep neural network (DNN) models that can be trained for domain-specific tasks. At the same time, large foundation models that are pretrai...

📖 Read original article


533. FusionRelight: Relighting Portraits in Real Time via Hybrid Domain Knowledge Fusion ​

Author: Qian Huang, Mayoore Selvarasa Jaiswal, Zhen Zhong, Rochelle Pereira, Jianyuan Min
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.LG

arXiv:2604.23094v2 Announce Type: replace-cross Abstract: Portrait relighting is a low-level vision problem in which physically plausible illumination transfer, identity preservation, and compact real-time inference must be considered together. Iterative diffusion-style methods can synthesize fine d...

📖 Read original article


534. ZenBrain: A Neuroscience-Inspired 7-Layer Memory Architecture for Autonomous AI Systems ​

Author: Alexander Bering
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2604.23878v3 Announce Type: replace-cross Abstract: ZenBrain is a seven-layer, neuroscience-derived memory architecture for LLM agents that unifies fifteen mechanisms - from Two-Factor synaptic consolidation to a Simulation-Selection sleep loop - under a single MemoryCoordinator: nine foundati...

📖 Read original article


535. Path-Lock Expert: Separating Reasoning Mode in Hybrid Thinking via Architecture-Level Separation ​

Author: Shouren Wang, Wang Yang, Chuang Ma, Debargha Ganguly, Vikash Singh, Chaoda Song, Xinpeng Li, Xianxuan Long, Vipin Chaudhary, Xiaotian Han
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.27201v3 Announce Type: replace-cross Abstract: Hybrid-thinking language models expose explicit /think and /no_think modes, but current designs do not separate them cleanly. Even in /no_think mode, models often emit long and self-reflective responses, causing reasoning leakage. Existing wo...

📖 Read original article


536. Multi-frame Restoration for 10 Hz Lissajous Confocal Laser Endomicroscopy ​

Author: Minhee Lee, Sangyoon Lee, Jiwook Lee, Minki Hong, Kyuyoung Kim, Won Hwa Kim, Jaeho Lee
Published: 8/11/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2605.00527v3 Announce Type: replace-cross Abstract: Lissajous confocal laser endomicroscopy (CLE) is a promising solution for high-speed in vivo optical biopsy for handheld scenarios. However, Lissajous scanning traces a resonant trajectory and samples only the visited pixels per frame; at hig...

📖 Read original article


537. Themis: Training Robust Multilingual Code Reward Models for Flexible Multi-Criteria Scoring ​

Author: Indraneil Paul, Goran Glava\v{s}, Iryna Gurevych
Published: 8/11/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2605.00754v4 Announce Type: replace-cross Abstract: Reward models (RMs) have become an indispensable fixture of the language model (LM) post-training playbook, enabling policy alignment and test-time scaling. Research on the application of RMs in code generation, however, has been comparativel...

📖 Read original article


538. Autonomous Reliability Qualification of Ga$_2$O$_3$-based diode sensors via Safe Active Learning ​

Author: Davi Febba, William A. Callahan, Anna Sacchi, Andriy Zakutayev
Published: 8/11/2026, 4:00:00 AM
Categories: physics.app-ph, cond-mat.mtrl-sci, cs.LG, cs.SY, eess.SY

arXiv:2605.00868v3 Announce Type: replace-cross Abstract: Ultra-wide bandgap (UWBG) Ga$_2$O$_3$ is a promising semiconductor for high-power and high-temperature electronics. Reliable qualification of these devices under extreme operating conditions is essential, yet conventional reliability testing ...

📖 Read original article


539. FitText: Evolving Agent Tool Ecologies via Memetic Retrieval ​

Author: Kyle Zheng, Han Zhang, Renliang Sun, Chenchen Ye, Wei Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.IR, cs.LG, cs.MA

arXiv:2605.02411v4 Announce Type: replace-cross Abstract: Efficient reasoning is not only a matter of shortening an answer trace; for tool-using agents, it also depends on whether the agent is reasoning over the right action space. As API ecosystems scale to tens of thousands of endpoints, the seman...

📖 Read original article


540. The Range Shrinks, the Threat Remains: Re-evaluating LLM Package Hallucinations on the 2026 Frontier-Model Cohort ​

Author: Aleksandr Churilov (Independent Researcher)
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.SE

arXiv:2605.17062v3 Announce Type: replace-cross Abstract: Spracklen et al. (USENIX Security '25) showed that code-generating large language models hallucinate package names that do not exist on PyPI or npm at rates ranging from 5.2% on commercial models to 21.7% on open-source models, creating an at...

📖 Read original article


541. Coupled Training with Privileged Information and Unlabeled Data ​

Author: Jiahao Shi, Omar Hagrass, Jason M. Klusowski
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2605.23268v2 Announce Type: replace-cross Abstract: In many prediction problems, we have extra information during training (for example, measurements that are expensive or slow to collect) that will not be available when the model is deployed. A common strategy is to first train a model that u...

📖 Read original article


542. Large language models reorganize representational geometry during in-context learning ​

Author: Hua-Dong Xiong, Li Ji-An, Robert C. Wilson, Kwonjoon Lee, Xue-Xin Wei
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, q-bio.NC

arXiv:2605.28854v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit remarkable flexibility in adapting to novel tasks from in-context examples without parameter updates, a capability known as in-context learning (ICL). Prior work has sought to understand this capability by...

📖 Read original article


543. Memory by Design: Probabilistic Sequence Layers ​

Author: Matthew Dowling, Hyungju Jeon, Cristina Savin, Il Memming Park
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2605.31163v2 Announce Type: replace-cross Abstract: We introduce the design-model framework: a way to derive efficient recurrent sequence maps from explicit assumptions about memory. A design model writes evidence into memory by exact Bayesian filtering; a query-dependent readout produces a pr...

📖 Read original article


544. The Deterministic Horizon: When Extended Reasoning Fails and Tool Delegation Becomes Necessary ​

Author: Dongxin Guo, Jikun Wu, Siu Ming Yiu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2606.00376v2 Announce Type: replace-cross Abstract: Extended chain-of-thought reasoning can degrade performance on deterministic state-tracking tasks, not solely because of preference biases but, on the evidence we present, because of information-theoretic limits in the capacity of decoder-onl...

📖 Read original article


545. OLIVE: Online Low-Rank Incremental Learning for Efficient Adaptive Exoskeletons ​

Author: Dong Liu, Yanxuan Yu, Ben Lengerich, Tong Geng, Ying Nian Wu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2606.05234v3 Announce Type: replace-cross Abstract: Wearable exoskeleton systems hold promise for restoring mobility in individuals with physical impairments, yet most existing controllers rely on static gait policies that cannot adapt to dynamic real-world environments or individual user char...

📖 Read original article


546. Persistent Priors, Preserved Targets: A Stroop-Style Paradigm for Lexical Override ​

Author: Han-yu Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2606.07555v5 Announce Type: replace-cross Abstract: Local definitions can assign a familiar word a temporary meaning while its usual associations remain useful elsewhere. We measure interference from those associations with a matched Stroop-style paradigm. A conflict prompt defines doctor as f...

📖 Read original article


547. Function-Vector Heads Are Two Populations: Writers and Cancellers in In-Context Learning ​

Author: Han-yu Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2606.07560v4 Announce Type: replace-cross Abstract: Function-vector (FV) analyses commonly identify attention heads by the magnitude of their causal contribution to in-context tasks. Magnitude does not retain the direction of the effect on the task readout. We preserve the sign and validate ca...

📖 Read original article


548. Capability Provenance in Language Models: A Case Study in Social Reasoning ​

Author: Glenn Matlin, Chandreyi Chakraborty, Saehee Eom, Mika Okamoto, Rayan Castilla, Louis Jaburi, Alvin Deng, Taywon Min, Lucia Quirke, Stella Biderman, Mark Riedl
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2606.19625v3 Announce Type: replace-cross Abstract: We use training-data attribution as an interpretable tool for capability discovery, mapping which regions of the pretraining corpus support social-reasoning versus STEM-reasoning in OLMo3-7B. Training-data attribution measures how strongly ea...

📖 Read original article


549. ORBIT: Training-Free Multi-Attribute Behavioral Steering via Orthogonal Subspace Rotation ​

Author: Narges Ghasemi, Amir Ziashahabi, Salman Avestimehr, Jonathan May
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.CV, cs.LG

arXiv:2606.22357v2 Announce Type: replace-cross Abstract: Language models are widely used in assistant settings, where controlling behavioral attributes is often essential. Activation steering modifies hidden-state representations at inference time, providing a lightweight, training-free mechanism t...

📖 Read original article


550. Unbiased Canonical Set-Valued Oracles Via Lattice Theory ​

Author: Jobst Heitzig
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.26418v3 Announce Type: replace-cross Abstract: An oracle that tells you the probability of some future event can change that very probability because you act on the answer. We argue that this performativity is OK as people consult oracles to be informed, and hence moved, by the answer. We...

📖 Read original article


551. Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction ​

Author: Chenguang Wang, Ming Li, Xinyue Zeng, Zhuochun Li, Hong Jiao, Tianyi Zhou, Dawei Zhou
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG

arXiv:2606.28186v3 Announce Type: replace-cross Abstract: Predicting human item difficulty is central to educational assessment, where reliable estimates support fairness and effective test construction. Existing methods often depend on costly human calibration or item-level textual representations,...

📖 Read original article


552. DataComp-VLM: Improved Open Datasets for Vision-Language Models ​

Author: Matteo Farina, Vishaal Udandarao, Thao Nguyen, Selim Kuzucu, Maximilian B"other, Andreas Hochlehnert, Adhiraj Ghosh, Marianna Nezhurina, Karsten Roth, Joschka Struber, Yuhui Zhang, Sebastian Dziadzio, Elaine Sui, Soumya Jahagirdar, Dhruba Ghosh, Hasan Hammoud, Thomas De Min, Simone Caldarella, Jehanzeb Mirza, Sedrick Keh, Mehdi Cherti, Hilde Kuehne, Bernt Schiele, Serena Yeung-Levy, Muhammad Ferjad Naeem, Federico Tombari, Ana Klimovic, Elisa Ricci, Matthias Bethge, Sewoong Oh, Ameya Prabhu, Alessio Tonioni, Jenia Jitsev, Massimiliano Mancini, Ludwig Schmidt, Nikhil Parthasarathy
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2606.28551v3 Announce Type: replace-cross Abstract: Building performant Vision-Language Models (VLMs) requires carefully curating large-scale training datasets, yet the community lacks systematic benchmarks for evaluating such curation strategies. We introduce DataComp for VLMs (DCVLM), a benc...

📖 Read original article


553. Message Passing Enables Efficient Reasoning ​

Author: Xuecheng Liu, Daman Arora, Gokul Swamy, Andrea Zanette
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.01077v2 Announce Type: replace-cross Abstract: While inference-time scaling has improved the reasoning abilities of large language models (LLMs), the need to generate long chains-of-thought (CoTs) is a computational bottleneck. Thus, in contrast to sequential scaling methods like CoT, rec...

📖 Read original article


554. The Moving Target: A Longitudinal Audit of Trust-Benchmark Score Drift Across Open-Source Chat LLM Release Lines ​

Author: Zhichao Fan, Yanhang Li, Zexin Zhuang, Xian Sun, Yingshuo Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2607.02587v2 Announce Type: replace-cross Abstract: Trust-benchmark scores reported on a chat-LLM release line are often carried across several checkpoints of the same line, as if the underlying model had not shifted between releases. We test that assumption. We audit four open-source release ...

📖 Read original article


555. UI-MOPD: Multi-Platform On-Policy Distillation for Unified GUI Agents ​

Author: Niu Lian, Tongbo Chen, Zhehao Yu, Chengzhen Duan, Fazhan Liu, Hui Liu, Pei Fu, Jian Luan, Heng Qu, Shu-Tao Xia, Jinpeng Wang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CV, cs.LG, cs.MM

arXiv:2607.04425v2 Announce Type: replace-cross Abstract: Recent advances in multimodal foundation models and agent systems have driven GUI agents from single-platform task execution toward cross-platform interaction. However, unified multi-platform GUI learning remains challenging: high-quality cro...

📖 Read original article


556. EHR-MPC: Inference-Time Control for Sepsis Treatment with Generative Patient Digital Twins ​

Author: Joshua Pickard, Wei Qi, Na Li, Ann Woolley, Lisa Cosimi, Roy Kishony, Deborah Hung
Published: 8/11/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, cs.SY, eess.SY, math.OC

arXiv:2607.08793v4 Announce Type: replace-cross Abstract: Sepsis is a leading cause of mortality, yet optimal treatment policies remain contested. Existing reinforcement learning (RL) approaches learn fixed strategies for sepsis treatment, limiting adaptability to changing clinical objectives during...

📖 Read original article


557. HIVE-3D: Hierarchical Voxel Enhancement for High-Quality 3D Scene Generation ​

Author: Bin Zang, Wenting Zheng, Xiaoliang Luo, Zhiyuan Fang, Shi Li, Lvchun Wang, Wei Yu, Yi Zhao, Tian Xie, Yuchi Huo, Rengan Xie
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.13468v2 Announce Type: replace-cross Abstract: Recently, a line of works can generate impressive 3D objects from a single image, but they are limited by restricted representation resolution, making them unsuitable for 3D scene generation. In this work, we introduce HIVE-3D, a novel method...

📖 Read original article


558. Concept-Guided Spatial Regularization for World Models in Atari Pong ​

Author: Yukuan Lu, Zaishuo Xia, Weyl Lu, Yubei Chen
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.15142v2 Announce Type: replace-cross Abstract: World models are usually evaluated as components of model-based reinforcement learning (MBRL) systems, leaving their standalone reliability understudied. We reproduce five visual world-model agents in Atari Pong -- DreamerV3, DIAMOND, TWISTER...

📖 Read original article


559. LLM-Driven AutoML for Cross-Lingual Handwritten OCR: Closed-Loop Neural Architecture Search with GPT-5, GPT-4o, and Claude Sonnet 4 ​

Author: Mobina Kashaniyan, Amirhossein Ghassemi, Nasser Mozayani
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.15509v2 Announce Type: replace-cross Abstract: We present a fully automated closed-loop AutoML framework that uses GPT-5, GPT-4o, and Claude Sonnet 4 as autonomous neural architecture designers for cross-lingual handwritten optical character recognition. Each large language model independ...

📖 Read original article


560. Exact Network Surgery: Functional Invariance and Gradient Plasticity in Reactive Computational Graphs ​

Author: Abdallah Khemais (ISITCOM, University of Sousse)
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.PL

arXiv:2607.16568v2 Announce Type: replace-cross Abstract: Function-preserving network growth techniques such as Net2Net and progressive stacking expand a model's capacity without destroying its learned function, but existing formulations either tolerate numerical perturbations or require a full rebu...

📖 Read original article


561. Attributes Should Come from Images, Not Class Names: Distribution-Conditioned Attribute Selection for Vision-Language Models ​

Author: Gautam Rajendrakumar Gare, Jia Shi, Zhiqiu Lin, Deepak Pathak, John Galeotti, Deva Ramanan
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV

arXiv:2607.18695v2 Announce Type: replace-cross Abstract: A popular route to interpretable zero-shot classification asks a large language model (LLM) to describe each class name and prompts CLIP with the resulting descriptors. We show that these descriptors carry little visual evidence of their own:...

📖 Read original article


562. Unified Static-Dynamic Pruning for Efficient LLM Inference ​

Author: Jinhyeok Kim, Yejoon Lee, Jaeyoung Do
Published: 8/11/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.AR, cs.LG

arXiv:2607.21985v2 Announce Type: replace-cross Abstract: The increasing deployment of large language models (LLMs) has magnified the computational and memory bottlenecks of autoregressive decoding, where low compute intensity and bandwidth-bound kernels dominate inference cost. Weight pruning offer...

📖 Read original article


563. When benchmark inferences do not compose: Projectibility in AI evaluation ​

Author: Brett Reynolds
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG

arXiv:2607.26159v2 Announce Type: replace-cross Abstract: An AI benchmark result rarely reaches a consequential claim in one step. Evaluators generalize it to further cases, interpret it as evidence of capability, extrapolate it to new tasks, transport it to another system or site, and combine it wi...

📖 Read original article


564. Attack Ensembles Expose a Safety-Utility Trade-off in Black-Box Guard Defenses Against Encoded VLM Jailbreaks ​

Author: Haoyu Zhang, Zhuoxi Wang, Shibo Zheng, Yi Feng, Xiao Luo, Zijian Xiao, Haowen Xu, Xiangchen Guan, Mohammad Zandsalimy, Shanu Sushmita
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2607.26574v2 Announce Type: replace-cross Abstract: Safety classifiers ("guards") are the dominant black-box defense for vision-language models, yet a guard judges an input's surface form, not its meaning: a harmful request re-encoded as set theory, formal logic, a classical language, code, or...

📖 Read original article


565. A Constitution-Grid Instrument for Data-Efficient RL Alignment (C-Guard) ​

Author: Xianling Zhang
Published: 8/11/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.00180v3 Announce Type: replace-cross Abstract: Conflicting objectives are general in RL alignment, and training on them data-efficiently is hard. Training a safety guard with RL means optimizing two objectives that conflict: catch real harm, and do not refuse benign prompts. Our finding i...

📖 Read original article


566. Evolutionary Curriculum Learning Improves Biological Sequence Modeling ​

Author: Richard Zhu, Kento Nishi
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-bio.BM, stat.ML

arXiv:2608.00697v2 Announce Type: replace-cross Abstract: Variational autoencoders (VAEs) trained on multiple sequence alignments (MSAs) have emerged as powerful generative models for biological sequences, with applications ranging from disease variant prediction to functional RNA design. However, s...

📖 Read original article


567. Emergence Invariance: From Symbolized Thought to Structural Control ​

Author: Yi Liu
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.01548v2 Announce Type: replace-cross Abstract: Language-first intelligence is constrained by which distinctions enter its symbolic record, which mappings its language--interpreter--environment complex can execute, and which possibilities can be realized with finite resources. We formalize...

📖 Read original article


568. Short-term load forecasting under EU-AI Act Requirements in Safety-Critical Environments: Results from a 41-day live challenge on the aggregated German transmission-grid load ​

Author: Thomas Bartz-Beielstein
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.05018v2 Announce Type: replace-cross Abstract: Short-term load forecasting (STLF) play a vital role in the electric power industry. It serves infrastructure that European and German law designate as critical. Determinism, reproducibility, and auditability are engineering requirements rath...

📖 Read original article


569. bioMoR: Biology-Guided Mixture-of-Recursions for Effective Genomic Learning ​

Author: Koushik Howlader, Tirtho Roy, Md Tauhidul Islam, Wei Le
Published: 8/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.06727v2 Announce Type: replace-cross Abstract: Transformer models for high-dimensional omics analysis process thousands of genes or pathways, although only a subset requires deep computation. Mixture-of-Recursions (MoR) improves efficiency through adaptive token-choice or expert-choice ro...

📖 Read original article


570. Establishing Boundary KKT Convergence of Mirror Descent through Reparameterization ​

Author: Kuangyu Ding, Kim-Chuan Toh
Published: 8/11/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2608.07248v2 Announce Type: replace-cross Abstract: Sequence convergence to a boundary Karush--Kuhn--Tucker (KKT) point has long remained unclear for nonconvex mirror descent with Legendre kernels. The difficulty arises from the blow-up of the gradient of the Legendre kernel at the boundary. R...

📖 Read original article