Skip to content

arXiv cs.LG - 2026-07-14 ​

488 items collected.


1. Knowledge Graphs Meet Graph Neural Networks: A Comprehensive Survey ​

Author: Chengcheng Sun, Jiayun Tian, Cheng Zhai, Zhixiao Wang, Yajie Song, Xiaobin Rui, Jian Zhang, Philip S. Yu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SI

arXiv:2607.09666v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have emerged as a powerful paradigm in Knowledge Graphs (KGs) due to their intrinsic ability to model graph-structured data. However, there remains a lack of a systematic review about GNN-based methodologies across the enti...

📖 Read original article


2. Position: Every Ground Truth is a Human Construction, not an Objective Truth ​

Author: Charlotte H"ogberg, Ericka Johnson, Kiri L. Wagstaff
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09668v1 Announce Type: new Abstract: Ground truth datasets play a fundamental role as reference values in the training and evaluation of machine learning models. This position paper argues that ground truths are not neutral objective measurements that are naturally given, but instead that...

📖 Read original article


3. AuditWeave: A Tamper-Evident, Auditor-Navigable Evidence Layer for AI-Assisted and Data-Transformation Workflows ​

Author: Vimal Nakrani
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.SE

arXiv:2607.09682v1 Announce Type: new Abstract: AI systems are increasingly used to assist consequential decisions in regulated domains such as auditing, finance, and healthcare. This creates a recurring obligation: an organization must be able to reconstruct, after the fact, which evidence informed...

📖 Read original article


4. Ablation, Statistical Inference, and Validation for KV-Cache Compression ​

Author: Paolo D'Alberto, Ashish Siarasao, Elliott Delaye, Rajeev Patwari
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT

arXiv:2607.09683v1 Announce Type: new Abstract: This study systematically compares Turbo-Quant and SpectralQuant KV-cache compression, evaluating non-dominated schemes, including WHT rotation with Beta Lloyd-Max and QJL, through a statistical validation methodology that separates systematic codec di...

📖 Read original article


5. SciML in the Wild: A Diagnostic Study of When Structural Priors Help and When They Hurt ​

Author: Vrishank Sai Anand, Prathamesh Dinesh Joshi, Raj Abhijit Dandekar, Rajat Dandekar, Sreedath Panat
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.09684v1 Announce Type: new Abstract: Scientific Machine Learning (SciML) methods such as Neural Ordinary Differential Equations (NODEs), Physics-Informed Neural Networks (PINNs), and Universal Differential Equations (UDEs) are most effective when structural priors reflect reliable governi...

📖 Read original article


6. MawForge: Memory-Bounded Expert Materialization for Local Mixture-of-Experts Inference ​

Author: Craig Opie
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09686v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) language models separate total parameter count from per-token active computation, but local inference systems often still require the full model, key-value cache, runtime buffers, and operatingsystem headroom to fit in f...

📖 Read original article


7. Prioritizing Search Space Regions in the Low Autocorrelation Binary Sequences Problem ​

Author: Bla\v{z} P\v{s}eni\v{c}nik, Borko Bo\v{s}kovi'c, Jan Popi'c, Janez Brest
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09688v1 Announce Type: new Abstract: Low autocorrelation binary sequences problem (LABS) is a hard combinatorial optimization challenge with important applications in communications, signal processing, and satellite navigation. This paper proposes a hybrid search framework that combines T...

📖 Read original article


8. What Context Does a Coding Agent Actually Need to Act? ​

Author: Brian Sam-Bodden
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.09691v1 Announce Type: new Abstract: A modern coding agent can hold an entire repository in its context window. Most of its reading is wasted -- and the interesting question is not how much context an agent can use, but what it actually \emph{needs}. We study that question at the moment i...

📖 Read original article


9. Reference-Based Distillation Detection in LLMs ​

Author: Rajat Rawat, Sizhe Chen, Akshay Anand, Michael Duan, Bob Rotsted, Sewon Min
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.09692v1 Announce Type: new Abstract: Model distillation -- training on outputs from stronger third-party models -- is widely used to boost performance, but raises concerns about unfair advantages and policy violations. This motivates a fundamental question: can we detect whether a model w...

📖 Read original article


10. Depth-Entropy Guided Sampling for Training-Free LLM Reasoning ​

Author: Zibin Meng, Peng Xie, Kani Chen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.09693v1 Announce Type: new Abstract: Reinforcement learning (RL) has become the dominant paradigm for improving the reasoning capabilities of large language models, but it requires expensive training, curated data, and reward signals. Recent work shows that sampling from sharpened base-mo...

📖 Read original article


11. Low-Rank Attention Residuals ​

Author: Jonathan Su
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.09694v1 Announce Type: new Abstract: Attention Residuals replace the fixed residual sum with depthwise attention over previous sub-layer outputs in large language models (LLMs), but use each output as both a full-dimensional key and value. This couples routing with representation and make...

📖 Read original article


12. FedCausal-Dyn: A Causal-Dynamic Paradigm for Federated Learning under Dynamic Feature Drift ​

Author: Kaijie Chen, Alex Johnson, Maria Garcia, Wei Zhang, Daniel Kim
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09695v1 Announce Type: new Abstract: This paper addresses the challenging problem of dynamic feature drift in federated learning, where data distributions evolve across clients and over time -- a common scenario in real-world applications like financial technology. Existing approaches oft...

📖 Read original article


13. Mitigating Early Training Collapse in CTR Models ​

Author: Ergun Bi\c{c}ici, Erkan \c{C}etinyama\c{c}
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.09696v1 Announce Type: new Abstract: Deep neural models for click-through rate prediction often exhibit a sharp decline in validation performance immediately after the first training epoch despite continued improvement in training loss. This instability restricts effective learning and li...

📖 Read original article


14. Safe responses matter: Output-aware safety guardrail mitigate over-refusal in MLLMs ​

Author: Jiayi Li, Kun Zhan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09697v1 Announce Type: new Abstract: Existing safety mechanisms for multimodal large language models (MLLMs) face a fundamental trade-off between safety and utility. Model fine-tuning achieves robust safety but compromises general utility. Input-side safety guardrails offer a lightweight ...

📖 Read original article


15. Quantum-Inspired Contextual Learning for Sparse-Ring Fraud Detection in Dynamic Transaction Graphs ​

Author: Behnam Tonekaboni, Hiroshi Yamauchi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09704v1 Announce Type: new Abstract: We present an exploratory benchmark and quantum-inspired modeling prototype for fraud screening in dynamic financial transaction graphs. Coordinated fraud may not be visible from individual transactions alone, but may emerge as a multi-period relationa...

📖 Read original article


16. Manifold Constrained Tabular Deep Neural Networks ​

Author: Tian Li, Lucy Robinson, Varun Ojha, Huizhi Liang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.09710v1 Announce Type: new Abstract: Tabular classification is often governed by local, condition-triggered rules rather than smooth global patterns. However, tabular deep neural networks (DNNs) are typically built upon Euclidean representations that favor smooth variations and semantic l...

📖 Read original article


17. EvoClawBench: Can Agents Learn Reusable Skills from Their Own Runs? ​

Author: Zhiyuan Peng, Xin Yin, Chenhao Ying, Zhe Cui, Zixiang Ding, Zhenhua Liu, Jiang Wu, Yuan Luo
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.SE

arXiv:2607.09711v1 Announce Type: new Abstract: Existing agent benchmarks primarily test task completion, tool use, or skill utility, but do not isolate whether a runtime can convert evidence from its own runs into reusable skills that improve fresh executions after authoring overhead. We introduce ...

📖 Read original article


18. ERP Data Provisioning Financial Control Testing ​

Author: Anitha Samudrala
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09712v1 Announce Type: new Abstract: Financial control testing increasingly depends on representative enterprise resource planning (ERP) data in quality environments, yet direct production copies expose personal, supplier, banking, and commercially sensitive records. This work presents Se...

📖 Read original article


19. Gauge dependence and structured-output corruption in sign-branched repetition penalties: measurements across models, inference stacks, and alternative repetition controls ​

Author: Peter Hollows
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.09791v1 Announce Type: new Abstract: The multiplicative repetition penalty shipped across the LLM inference ecosystem (HuggingFace, vLLM, llama.cpp, and a dozen further engines) branches on the sign of each raw logit (divide positives by theta, multiply negatives). But the softmax is unch...

📖 Read original article


20. Metadata-Free Meta-Reweighted Direct Preference Optimization under Noisy Preference Labels ​

Author: Hua Qu, Yifan Li, Xiaodong Yuan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09796v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) has become an important method for aligning large language models (LLMs) with human preferences because it removes the need for explicit reward modeling and reinforcement learning optimization. However, its performa...

📖 Read original article


21. JEPA for AI-Native 6G: Predictive Representations and Open Challenges ​

Author: Sheikh Salman Hassan, Irshad A. Meer, Almoatssimbillah Saifaldawla, Yan Kyaw Tun, Mustafa Ozger, Madyan Alsenwi, Nguyen Van Huynh, Woong-Hee Lee, Cedomir Stefanovic, Mathini Sellathurai, Henk Wymeersch, Tharmalingam Ratnarajah
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NI

arXiv:2607.09798v1 Announce Type: new Abstract: Sixth-generation (6G) networks are moving toward AI-native operation, where learning modules are embedded across the radio access network (RAN), edge, and core. This transition requires learning from limited labels, heterogeneous wireless and network d...

📖 Read original article


22. The Silent Freeze: Predicting When Low-Precision Training Stops Learning ​

Author: Zekai Shang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09800v1 Announce Type: new Abstract: Training in reduced floating-point precision can silently halt learning: when a gradient-descent weight update falls below half the unit in the last place (ULP) of the weight, it rounds away and that coordinate freezes while its gradient is still nonze...

📖 Read original article


23. Discovering Latent Response Laws in Forced Physical Systems ​

Author: Yi Zhu, Su Chen, Xiaojun Li, Xiuli Du
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09801v1 Announce Type: new Abstract: Governing equations provide compact descriptions of physical systems, yet the variables in which they are simple are often hidden in high-dimensional measurements. This challenge is sharper for forced systems, whose responses depend on both intrinsic d...

📖 Read original article


24. Quota Marketplace: Dynamic Pricing for Efficient Allocation of ML Training Resources ​

Author: Balasubramanian Sivan, Renato Paes Leme, Mihai Tiuca, Ian McFarlane, Vasilis Gkatzelis, Nehal Mehta, Soheil Hassas Yeganeh, Vahab Mirrokni, Amin Vahdat
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.GT

arXiv:2607.09802v1 Announce Type: new Abstract: The escalating demand for Machine Learning (ML) training resources in recent years has resulted in a substantial gap between the high demand and the available supply. Efficient allocation of these scarce and expensive resources is crucial for organizat...

📖 Read original article


25. Spectral Origins of the Self-Correction Blind Spot in Autoregressive Generation ​

Author: Ingrid Petrova, Luan Vejsiu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.09803v1 Announce Type: new Abstract: Large autoregressive language models exhibit a self-correction blind spot: they reliably fix identical errors when attributed to an external source yet fail to fix the same errors in their own outputs. Prior work has documented this phenomenon empirica...

📖 Read original article


26. RUBRIC: Realism--Utility Balanced Ranking for Imbalanced Classification ​

Author: Yanxuan Yu, Dong liu, Renata Borovica-Gajic, Ying Nian Wu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09816v1 Announce Type: new Abstract: Class imbalance poses a fundamental challenge in risk-sensitive applications such as fraud detection and medical diagnosis, where minority-class samples are scarce yet critical for accurate classification. Existing oversampling methods generate synthet...

📖 Read original article


27. Estimation, Prediction, and Assortment Optimization for Markov Chain Choice Models with Panel Data ​

Author: Yalcin Akcay, Gerardo Berbeglia, Young-San Lin
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.09817v1 Announce Type: new Abstract: We propose a framework for the Markov chain (MC) choice model with panel data, including parameter estimation, personalized choice prediction, and personalized assortment optimization. In contrast to the traditional setting, which assumes that each tra...

📖 Read original article


28. Learning Predictive Ambiguity Sets for Decision-Focused Distributionally Robust Optimization ​

Author: Junjie Guo
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP

arXiv:2607.09820v1 Announce Type: new Abstract: Predict-then-optimize systems usually compress uncertainty into a point forecast and then solve a downstream optimization problem as if the forecast were reliable. Distributionally robust optimization (DRO) offers protection against misspecification, b...

📖 Read original article


29. A Strong Balanced-Softmax Classifier-Retraining Baseline for Long-Tailed Recognition ​

Author: Juan Terven, Diana Margarita C'ordova Esparza, Julio Alejandro Romero Gonzalez, Edgar Arturo Ch'avez Urbiola, Francisco Javier Willars Rodriguez, Juan Bautista Hurtado Ramos, Alfonso Ramirez Pedraza
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.09832v1 Announce Type: new Abstract: Long-tailed recognition methods often modify losses, margins, or representations to reduce the dominance of frequent classes. We ask whether, after Balanced Softmax training, the remaining tail error can be reduced by retraining only the classifier. We...

📖 Read original article


30. From Direction to Magnitude: How Multimodal Instruction-Tuning Reorganizes the Geometric Encoding of Identity-Specifying Prompts in Transformer Hidden States ​

Author: Jorge A. Castillo, Marco Torres Y'evenes, Juan Carlos Lanas
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.09842v1 Announce Type: new Abstract: We investigate whether identity-specifying system prompts produce statistically distinguishable geometric fingerprints in the hidden-state trajectories of four open-weight transformer language models spanning four post-training regimes: no training (Ge...

📖 Read original article


31. Nonlinear Axiomatic Attribution for Cooperative Games ​

Author: Weida Li, Zhuanghua Liu, Yaoliang Yu, Bryan Kian Hsiang Low
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09869v1 Announce Type: new Abstract: The Shapley value is a widely used concept in attribution problems, as it uniquely satisfies the axioms of linearity, consistency, equal treatment, and efficiency. Often, the inclusion AUC metric is used to evaluate the quality of player rankings, in o...

📖 Read original article


32. Serving the Long Tail: Training-Free LLM Candidate Generation for Vacation Rental Marketplaces ​

Author: Syed Mohammed Arshad Zaidi, Eric Rincon, Shayan Hassantabar
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2607.09877v1 Announce Type: new Abstract: Vacation rental marketplaces face a structural imbalance on the supply side: a small fraction of properties receive most user interactions, while the long tail of new, niche, and seasonal listings generates too little behavioral signal for collaborativ...

📖 Read original article


33. Nonparametric Bayesian Inverse Reinforcement Learning with Data-Parallel Gibbs Sampling ​

Author: Sai Anirudh Katupilla, Shreeya Dasa Lakshminath
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09886v1 Announce Type: new Abstract: Inverse Reinforcement Learning recovers reward functions from expert demonstrations, but standard formulations assume that all demonstrations come from a single expert. When demonstrations are pooled from multiple experts with distinct preferences, par...

📖 Read original article


34. Remembering Distinct Items, Not Tokens: A Learnable Dirichlet-Process Cache Between State-Space Models and Attention ​

Author: Siddharth Pal, Viktoria Rojkova
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.NE

arXiv:2607.09889v1 Announce Type: new Abstract: Fixed-state sequence models compress an unbounded past into a bounded state, which caps their associative recall at roughly the state dimension; attention escapes the cap by keeping a key-value entry for every token, at quadratic compute and a cache th...

📖 Read original article


35. SMETA-ZSL:Semantic Meta-Alignment for Zero-Shot Threat Classification ​

Author: Ivan Alejandro Montoya Sanchez, Anantaa Kotal, Aritran Piplai
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2607.09936v1 Announce Type: new Abstract: Cybersecurity systems must adapt rapidly to emerging threats. However, labeled data for new threat categories is unavailable when those threats first appear. Generalized zero-shot learning offers a natural solution by enabling recognition of unseen cla...

📖 Read original article


36. A Foundation Model for Multimodal Event Sequences in Financial Applications ​

Author: Nikita Rusakov, Vladislav Meshkov, Konstantin Zorin, Gleb Zaripov, Alexander Uglov, Alexey Vasilev, Anton Klenitskiy
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.09955v1 Announce Type: new Abstract: Predictive modeling is a core component of modern financial services, where a wide range of tasks are traditionally addressed using separate models trained on manually engineered tabular features. This task-specific approach limits reuse and makes it d...

📖 Read original article


37. Optimizing ARDL Models for Retail Sales Forecasting and Fair Pricing ​

Author: Sujay Uday Rittikar
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CE

arXiv:2607.09956v1 Announce Type: new Abstract: Pricing food products to balance profitability with consumer welfare is a central challenge for retailers. Dynamic pricing is widely used to maximize revenue, yet most pricing models optimize business objectives while overlooking consumer fairness. Thi...

📖 Read original article


38. Learning in Curved Weight Space:Exponential-Linear Weight Reparameterization for Improved Optimization ​

Author: Ethan Smith
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.09967v2 Announce Type: new Abstract: Many neural networks operations have a multiplicative nature rather than additive: halving or doubling a norm are analogous relatively but require unequal optimization distances when taking linear steps. Adaptive optimizers such as Adam normalize updat...

📖 Read original article


39. Multimodal Routing for Interpretable, Robust, and Auditable Clinical Prediction ​

Author: Nikkie Hooman, Zhongjie Wu, Eric C. Larson, Mehak Gupta
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09982v1 Announce Type: new Abstract: Electronic health record (EHR) data are inherently multimodal, and leveraging multiple modalities can improve predictive performance. However, most existing approaches rely on deep fusion, which obscures how individual modalities contribute to predicti...

📖 Read original article


40. Vilya-1: An all-atom foundation model for macrocycle structure prediction and design ​

Author: Vilya Research, :, Pascal Sturmfels, Milad Salem, Naozumi Hiranuma, Stephen Rettie, Xiaoliang Pan, Benjamin D. Sellers, Adam P. Moyer, Patrick J. Salveson, Ivan Anishchanka
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM

arXiv:2607.09998v1 Announce Type: new Abstract: Macrocyclic peptides are an increasingly important therapeutic modality, but existing computational methods for modeling their structures and properties are limited in scope and do not generalize well across the synthetically accessible chemical space....

📖 Read original article


41. MLPs are Hebbians: Constructing Efficient Fact-Storing MLPs for Transformers ​

Author: Roberto Garcia, Jerry Liu, Ronny Junkins, Sabri Eyuboglu, Atri Rudra, Christopher R'e
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10034v1 Announce Type: new Abstract: Large language models (LLMs) store factual knowledge in their parameters. While recent work has shown that this knowledge resides in MLP layers, existing constructive and mechanistic interpretability models of fact-storage in LLMs fail to explain the s...

📖 Read original article


42. FlashTrie: A GPU-Accelerated Constrained Beam Search for Generative Retrieval ​

Author: Dakshitha Anandakumar, Anurag Mukkara, Wenxiang Hu, Jiusheng Chen, M Akash Kumar, Ting Ye, Qiang Lou, Jian Jiao
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10044v1 Announce Type: new Abstract: Constrained decoding is essential in generative retrieval, where document identifiers generated directly from a query must exactly match a predefined library of valid IDs. At scale, decoding is often constrained using a trie with beam search but most i...

📖 Read original article


43. Conservation Laws for Diffusion Models ​

Author: Ziv Aharoni, Henry D. Pfister
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, stat.ML

arXiv:2607.10067v1 Announce Type: new Abstract: While autoregressive models optimize the exact data likelihood via the chain rule, diffusion models are typically trained with denoising objectives. We develop conservation laws based on generalized extrinsic information transfer (GEXIT) functions for ...

📖 Read original article


44. Error Aware Distribution Prediction for Lightweight Implicit Neural Representations ​

Author: Zhimin Li, Jake D. Balla, Joshua A. Levine
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.GR

arXiv:2607.10068v1 Announce Type: new Abstract: Implicit neural representations (INRs) offer compact encoding of volumes, but as lossy approximators, inevitably have prediction errors. We consider INRs that can simultaneously encode relative error scales by predicting distributions using tools from ...

📖 Read original article


45. Distance-Preserving Embeddings in Inhomogeneous Random Graphs ​

Author: My Le, Luana Ruiz, Souvik Dhara
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10074v1 Announce Type: new Abstract: Graph machine learning provides powerful tools for understanding complex networks and learning meaningful node representations. A central challenge, however, is designing embeddings with minimal distortion of both local and global functionals, such as ...

📖 Read original article


46. TabLoRA: Parameter-Efficient Low-Rank Ensemble Learning for Large-Scale Tabular Data ​

Author: Jiaqi Luo, Shixin Xu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10077v1 Announce Type: new Abstract: Tabular learning is still dominated by gradient-boosted decision trees (GBDTs), while recent deep learning approaches have become increasingly competitive. However, applying deep tabular models to large-scale datasets remains challenging, as large samp...

📖 Read original article


47. When Data Imbalance Helps: Robust Generalization Through Shortcut Saturation ​

Author: Cheng-Ting Chou, Duc Binh Hoang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10116v1 Announce Type: new Abstract: We study robust generalization under spurious correlations: tasks where a shortcut feature is correlated with the true label in training but anti-correlated in an adversarial held-out split. Varying the spurious ratio $r$ (the fraction of training exam...

📖 Read original article


48. GAE: Graph-Augmented Evolution for Scientific Discovery via Reinforcement Optimization ​

Author: Xuanzhou Chen, Taoli Cheng
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10127v1 Announce Type: new Abstract: Evolutionary program search guided by Large Language Models (LLMs) has emerged as a powerful paradigm for automated scientific discovery. However, current approaches are fundamentally constrained by three bottlenecks: structurally blind parent selectio...

📖 Read original article


49. Energy-guided Recursive Model ​

Author: Yifei Zhao, Ying Tang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.10128v1 Announce Type: new Abstract: Recursive reasoning models address structured problems by repeatedly updating latent states of small neural networks. However, their test-time scaling lacks a principled inference mechanism: increasing depth or stochastic breadth generates more traject...

📖 Read original article


50. SALT-GNN: Handling Dense Neighborhoods in Anti-Money Laundering Graphs via Statistics-Aware Attention ​

Author: Lidia Losavio, Francesco Sovrano, Dario Fenoglio, Martin Gjoreski, Marc Langheinrich
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10131v1 Announce Type: new Abstract: Money laundering threatens financial stability and exposes institutions to penalties, motivating automated detection. Because laundering schemes often emerge through relational patterns, graph neural networks (GNNs) are increasingly used for anti-money...

📖 Read original article


51. LeRoPE: Learnable RoPE Frequencies Improve Language Modeling ​

Author: Petros Karypis, Sean O'Brien, Shreyas Kadekodi, Rui Zhu, Julian McAuley
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10134v1 Announce Type: new Abstract: Rotary Positional Encodings (RoPE) are currently the most popular positional encodings used in modern language models. RoPE rotates two-dimensional chunks of query and key vectors, operating as a function of their relative positional offset. The positi...

📖 Read original article


52. RDQ: Residual Distribution Quantization for Large Language Models ​

Author: Prateek Singh
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10137v1 Announce Type: new Abstract: Post-training quantization (PTQ) of large language models degrades sharply below 4-bit precision. We identify the root cause as residual stream distributional drift: quantization noise injected at each transformer layer accumulates in the shared residu...

📖 Read original article


53. LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning ​

Author: Ning Liu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10139v1 Announce Type: new Abstract: Selecting the correct answer from a pool of candidate reasoning chains is the engine of test-time scaling, yet the standard selectors each carry a cost: self-consistency inherits the errors of the single model it resamples, and trained reward models ne...

📖 Read original article


54. Beyond Euclidean Clipping: Overcoming Exploration Collapse in LLM RL via Riemannian Isometric Policy Optimization ​

Author: Zhicheng Cai, Xinyuan Guo, Hanlin Wu, Mingxuan Wang, Wei-Ying Ma, Ya-Qin Zhang, Hao Zhou
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10169v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a dominant paradigm for enhancing LLMs' reasoning capabilities. However, RL algorithms with PPO-Clip are inherently limited by exploration collapse. Subsequent works remain primarily heuristic and fail to identify...

📖 Read original article


55. Knowledge-Conditioned, Single-Pass LLM Synthesis of Executable Unity Game Scenes: A Compiler Error Census across 26 Goal Playable Concepts ​

Author: Hugh Xuechen Liu, K{\i}van\c{c} Tatar
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.SE

arXiv:2607.10187v1 Announce Type: new Abstract: Large language models (LLMs) write Unity C# for game scenes. Yet nearly all demonstrations rest on an iterative repair loop that regenerates code until it compiles, conflating what the model writes with what the loop fixes. We remove the loop and eval...

📖 Read original article


56. PhysMRV: Physical Memory Retrieval and Verification for Physics Plausibility Reasoning ​

Author: Wenyuan Wang, Lianyu Hu, Hao Wang, Yang Liu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.10190v1 Announce Type: new Abstract: Video-language models (VLMs) have achieved remarkable performance on video understanding and visual question answering, yet they remain unreliable in reasoning about physical plausibility, where understanding object interactions, causal dynamics, and f...

📖 Read original article


57. Generative Augmentation of Raman Spectra for Glioma Classification ​

Author: Andrei Iu\c{s}an, Iulian Vasile, Daria Voiculescu, Ion Petre, Andrei P\u{a}un, Bogdan Oancea, Mihaela P\u{a}un
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10196v1 Announce Type: new Abstract: Access to sufficiently large biomedical datasets remains a major obstacle for machine learning in Raman spectroscopy-based diagnostics. In particular, for glioma analysis, datasets are typically small and heterogeneous, affected by acquisition-specific...

📖 Read original article


58. The Differential Neural Tangent Kernel and Its Positivity ​

Author: Bangti Jin, Longjun Wu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, stat.ML

arXiv:2607.10200v1 Announce Type: new Abstract: The Neural Tangent Kernel (NTK) is one powerful tool for analyzing the training dynamics of neural networks in the over-parameterized regime. Recently, the theoretical framework has been extended to physics-informed neural networks (PINNs) for solving ...

📖 Read original article


59. Two Confounds in Cross-Model Value Comparison: Response Determinism and the Access Harness ​

Author: Hong-In Won, Jinseok Jang, Hyoseop Kim
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10202v1 Announce Type: new Abstract: Cross-model comparisons read divergence in value dispositions as evidence that language models hold individuated values. Under single-draw measurement this conflates two quantities: a difference in central tendency (a genuine value difference) and a di...

📖 Read original article


60. Adaptive Compute in Latent World Models: When Depth Helps, Hurts, or Doesn't Matter ​

Author: Achyuthan Sivasankar
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10203v2 Announce Type: new Abstract: Adaptive-compute world models -- early-exit or mixture-of-depths predictors that spend variable depth per step -- assume depth buys better predictions and can be routed adaptively. In autoregressive rollouts, the first assumption requires depth's per-s...

📖 Read original article


61. Exploratory Analysis of Deep Learning Models for Forecasting Meteorological Parameters in the Agricultural Sector ​

Author: Piotr Sikora, Sotirios Kontogiannis
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10208v1 Announce Type: new Abstract: Accurate meteorological forecasting is essential for agricultural planning, irrigation management, and environmental decision support. This study conducts a comparative evaluation of recurrent and hybrid deep learning architectures for multivariate for...

📖 Read original article


62. DSSMs: State Space Models with Explicit Memory via Delay Differential Equations ​

Author: Yixiao Qian, Song Chen, Jiaxu Liu, Shengze Cai, Chao Xu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10244v1 Announce Type: new Abstract: State Space Models (SSMs) have emerged as a powerful paradigm for efficient long-sequence modeling, offering parallel training and fast linear-time recurrent inference. However, like other recurrent architectures, SSMs must compress an unbounded histor...

📖 Read original article


63. Data-Driven Telecom Marketing Optimization: A Machine Learning-Based Churn Prediction and Customer Segmentation Framework ​

Author: Nada Ali, Lina Ahmed, Tahani Abdalla Attia Gasmalla
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10260v1 Announce Type: new Abstract: Customer churn is a major challenge for telecommunication companies, directly eroding revenue and long term customer relationships. Traditional retention programs rely on generic, not personalized incentives and lack the precision to identify high risk...

📖 Read original article


64. Sharper Analysis of Single-Loop Methods for Bilevel Optimization ​

Author: Yubo Zhou, Jun Shu, Luo Luo, Junmin Liu, Deyu Meng, Guang Dai, Haishan Ye
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10263v1 Announce Type: new Abstract: Bilevel optimization underpins many machine learning applications, including hyperparameter optimization, meta-learning, neural architecture search, and reinforcement learning. While hypergradient-based methods have advanced significantly, a gap persis...

📖 Read original article


65. Interpreting learning dynamics of autoencoders: Transient scaling and emerging concepts of the Ising model ​

Author: Max Weinmann, Miriam Klopotek
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, cond-mat.stat-mech

arXiv:2607.10285v1 Announce Type: new Abstract: We study how unsupervised autoencoders trained on microscopic spin configurations from the Ising model learn macroscopic, theory-relevant variables underlying the data-generating process. Without embedding domain knowledge, we mimic a typical discovery...

📖 Read original article


66. Empowering Long-form Omni-modal Understanding with Robust Audio Perception ​

Author: Kaiying Yan, Luoyi Sun, Xiao Zhou, Weidi Xie
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10299v1 Announce Type: new Abstract: Recent advances in large-scale multimodal models have drivenremarkable progress in vision-language tasks; however, comprehensiveomni-modal understanding remains under-explored, largely due to thescarcity of datasets with rich, explicitly aligned audito...

📖 Read original article


67. A Control Theory of Predictability in Latent World Models ​

Author: Hanzhe You, Yonggang Zhang, Maohao Ran, Zhiqin Yang, Zhenyuan Zhang, Wei Xue, Jun Song, Xinmei Tian, Yike Guo
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10362v1 Announce Type: new Abstract: Latent world models are trained to predict future states in a learned representation and are then deployed inside a planner that selects actions by simulating them forward. Current practice adopts the prediction error, the single- or multi-step rollout...

📖 Read original article


68. A Hyperbolic Neural Closure for M1 Radiation Transfer ​

Author: Bongseok Kim, Jiahao Zhang, Johannes Krotz, Dinshaw Balsara, Ryan McClarren, Guang Lin
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE, cs.NA, math.NA

arXiv:2607.10364v1 Announce Type: new Abstract: In radiation transfer simulations, an M1 method achieves substantial computational savings by replacing the full angular transport equation with a low-order moment system. Because this reduced system is not closed, a closure model is required to repres...

📖 Read original article


69. Machine Learning-based Correlation of Charpy Impact Properties Between Sub-sized and Standard-sized Specimens for Nuclear Structural Materials ​

Author: Yugandhar Kasala Sreenivasulu, Isshu Lee, John W. Merickel, Fei Xu, Yalei Tang, Joshua E. Rittenhouse, Aleksandar Vakanski, Rongjie Song
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10412v1 Announce Type: new Abstract: Reliable correlations of Charpy impact test results between sub-sized and full-sized specimens are essential for structural integrity assessments, particularly in nuclear applications, where spatial constraints and limited material volume restrict spec...

📖 Read original article


70. Context by Distinct Information: An Auditable Dirichlet-Process Working Memory for Long, Redundant Context Streams ​

Author: Siddharth Pal, Viktoria Rojkova
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.IR

arXiv:2607.10441v1 Announce Type: new Abstract: Context engineering decides what information a model carries forward, and current designs meter it in tokens: compressing the past into a bounded recurrent state, keeping a key-value entry for every token, or imposing a fixed budget through a window or...

📖 Read original article


71. Pitfalls of Administrative Censoring in Survival Models with Time-Indexed Inputs ​

Author: Yanqi Xu, Hui Dai, Carlos Fernandez-Granda, Krzysztof J. Geras, Yiqiu Shen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.AP

arXiv:2607.10466v1 Announce Type: new Abstract: Survival models can model time-to-event outcomes using partially observed data. They are widely used in clinical prediction, including cancer risk, disease progression, treatment response, and mortality. Recent models often rely on rich inputs collecte...

📖 Read original article


72. Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards ​

Author: Pengfei Cai, Utkarsh Utkarsh, Alan Edelman, Christopher Vincent Rackauckas, Rafael Gomez-Bombarelli
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE

arXiv:2607.10474v1 Announce Type: new Abstract: Partial differential equations (PDEs) are foundational to modeling in science and engineering, but constructing reliable numerical solvers remains labor-intensive, demanding expert knowledge of discretization schemes, stability conditions, and boundary...

📖 Read original article


73. ARMOR: Stabilizing On-Policy LLM RL with Off-Policy Anchor Samples ​

Author: Kexin Huang, Junkang Wu, Jinda Lu, Shuo Yang, Chiyu Ma, Jiancan Wu, Xiang Wang, Xiangnan He, Guoyin Wang, Jingren Zhou
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.10481v1 Announce Type: new Abstract: Reinforcement learning (RL) has significantly enhanced the reasoning capabilities of large language models (LLMs), yet the training process remains notoriously fragile. In this work, we investigate a critical source of this instability: over-optimizati...

📖 Read original article


74. EvidentialRAG: Quantifying and Mitigating Information Conflict in Multi-Source Retrieval-Augmented Generation via Evidential Deep Learning ​

Author: S M Asif Hossain, Ruksat Khan Shayoni, M. F. Mridha
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10491v1 Announce Type: new Abstract: Retrieval-augmented generation grounds large language models in external evidence, but most pipelines still treat retrieved passages as deterministic and mutually consistent context. In open information environments, retrieved sources may disagree beca...

📖 Read original article


75. Learning from Noise: Effective-Rank Collapse and Out-of-Distribution Rejection in Restricted Boltzmann Machines ​

Author: Oshada Rathnayake, Nikhil Shukla
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn

arXiv:2607.10506v1 Announce Type: new Abstract: Restricted Boltzmann machines (RBMs) represent data by shaping an energy landscape over visible and hidden configurations, but their discriminative use is fragile under out-of-distribution (OOD) inputs: samples outside the training distribution can be ...

📖 Read original article


76. Conditional Optimal Bridge for Riemannian Activation Steering ​

Author: Seyed Arshan Dalili, Ajay Narayanan Sridhar, Vijaykrishnan Narayanan, Mehrdad Mahdavi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10517v1 Announce Type: new Abstract: Activation steering offers a lightweight alternative to fine-tuning for controlling large language models at inference time. While many existing methods implicitly optimize a log-density-ratio objective between desired and undesired activation distribu...

📖 Read original article


77. LLM-PDESR: Robust PDE Discovery via Subdomain Weighted Residuals and LLM-Guided Symbolic Hypothesis Generation ​

Author: Jinyang Du, Hao Ma, Xiaohu Shi, Bo Yang, Yanchun Liang, Heow Pueh Lee, Chunguo Wu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10546v1 Announce Type: new Abstract: Discovering governing partial differential equations (PDEs) from noisy observational data is a fundamental challenge in scientific machine learning. Traditional symbolic regression (SR) methods often struggle to identify accurate equations within vast ...

📖 Read original article


78. Learning from Local Walks on Dynamic Graphs with Bandit Feedback ​

Author: Sourav Chakraborty, Amit Kiran Rege, Claire Monteleoni, Lijun Chen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.10571v1 Announce Type: new Abstract: We study stochastic multi-armed bandits on dynamic graphs, where arms correspond to the vertices of a network with time-varying edges. In this setting, the learner is restricted to local movement, selecting only its current node or an immediate neighbo...

📖 Read original article


79. MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference ​

Author: Venkatesha Matam, Keon Kim
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10582v1 Announce Type: new Abstract: Large language model (LLM) agents accumulate heterogeneous context, including system instructions, plans, user turns, retrieved documents, tool outputs, and intermediate reasoning, whose key-value (KV) cache can become a major memory bottleneck. Existi...

📖 Read original article


80. Sharp Concentration Bounds for Bundle-Valued Statistics on Manifolds ​

Author: Swagatam Das, Vaclav Snasel
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10592v1 Announce Type: new Abstract: Many geometric statistics and manifold learning pipelines routinely produce observations -- such as tangent vectors or local frames -- whose natural home is a varying family of fibers attached to different points of a base manifold, rather than a singl...

📖 Read original article


81. AutoNorm: Understanding Adaptive Normalization in Transformers through Differentiable Gating ​

Author: Piyush Kaushik Bhattacharyya, Divyanshu Rai, Swastik Singh, Kumar Aakash, Ayush Ranjan, Krutika Verma
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10593v1 Announce Type: new Abstract: Normalization is a critical component for stabilizing Transformer training, yet the choice between static strategies such as Layer Normalization (LN) and adaptive alternatives remains largely task-dependent. In this paper, we investigate a key optimiza...

📖 Read original article


82. M+Adam: Low-Precision Training via Additive-Multiplicative Optimization ​

Author: Xiaoyuan Liang, Sebastian Loeschcke, Mads Toftrup, Anima Anandkumar
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10611v1 Announce Type: new Abstract: Training with quantized weights can reduce costs but often results in degraded accuracy, especially when optimization is carried out in low precision, without storing high-precision copies. We identify a key failure mode: under low precision, standard ...

📖 Read original article


83. modelDNA: Calibrated Lineage Verification and Merge Decomposition from Sampled Weight Fingerprints ​

Author: Muhammad Awais Bin Adil, Saad Aamir
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10617v1 Announce Type: new Abstract: The lineage graph of open-weight language models is self-reported: Hugging Face's base_model metadata field is optional and unverified, and over 60% of Hub models document no parentage at all. Methods for detecting lineage from weights exist in the res...

📖 Read original article


84. Auditing Construct Overlap in Explainable Machine Learning: Evidence from Burnout-Depression Prediction Across Student Cohorts ​

Author: Alireza Dehghan, Negin Ashrafi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10633v1 Announce Type: new Abstract: Explainable machine learning (XML) pipelines applied to composite mental health outcomes can produce apparently-robust, cross-population-stable risk hierarchies that are largely artefacts of how the outcome was constructed. We demonstrate this using an...

📖 Read original article


85. Modernizing HEBO: a robust Bayesian optimization baseline for practical heteroskedastic and non-stationary problems ​

Author: L. A. Zhukov, E. V. Shaburova, D. V. Antonets
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10669v1 Announce Type: new Abstract: Bayesian optimization is increasingly used to guide data-efficient experimentation in chemistry, materials science, and related laboratory settings, but its practical performance depends strongly on how well surrogate-model assumptions match the geomet...

📖 Read original article


86. From Self-Attention to Connection Laplacian: A Unified Operator View of Transformers ​

Author: Binbin Lin, Wei Chen, Yalun Li, Wenxiao Wang, Jieping Ye, Xiaofei He
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.10677v1 Announce Type: new Abstract: Self-attention is a ubiquitous primitive in modern sequence models, yet its operator-level geometry is only partially understood. We view a token sequence as a vector field over the token-position graph and identify attention as a connection walk: mess...

📖 Read original article


87. LayerNorm as Implicit Gain Control in Looped Transformers ​

Author: Matthias M. M. Buehlmaier
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2607.10681v1 Announce Type: new Abstract: In pre-LayerNorm looped transformers, LayerNorm inside the recurrent block acts as an implicit gain controller: by coupling the block's local Lipschitz constant inversely to the activation scale, it renders the recurrence Jacobian non-normal -- asympto...

📖 Read original article


88. Learning to Fine-tune Foundation Models under Resource Limitations ​

Author: Thomas Tsouparopoulos, Iordanis Koutsopoulos
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10694v1 Announce Type: new Abstract: We study the problem of optimal continual fine-tuning for a pre-trained Foundation Model deployed at a resource-limited device. At each time slot, a new batch of training data arrives, and the controller is faced with two options: either use the data t...

📖 Read original article


89. On the modality gap and the contrastive loss in multi-modal representation learning ​

Author: Fabian Mager, Hiba Nassar, Lars Kai Hansen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.10698v1 Announce Type: new Abstract: We study the modality gap in CLIP-style dual-encoder contrastive learning, where image and text embeddings remain misaligned despite being trained in a shared space. We argue that the gap is induced by a failure of the InfoNCE formulation with independ...

📖 Read original article


90. Scaffold splits hide structural-frontier failures in ADMET models ​

Author: Jiacheng Zheng, Chang Guo, Zixuan Wang, Xinyu Liu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM

arXiv:2607.10729v1 Announce Type: new Abstract: Molecular property models are commonly evaluated by holding out Bemis--Murcko scaffolds, yet a scaffold identifier is only one notion of chemical unfamiliarity. We introduce a label-free structural-frontier split that reserves the sparsest and most phy...

📖 Read original article


91. To Answer or to Abstain: Mitigating Search-Agent Hallucinations via Abstention-Aware Reinforcement Learning ​

Author: Fengji Zhang, Tianyu Fan, Yuxiang Zheng, Xinyao Niu, Chengen Huang, Jacky Keung, Bei Chen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.10738v1 Announce Type: new Abstract: Recent advances in equipping Large Language Models (LLMs) with search tools and outcome-reward reinforcement learning (RL) have achieved new state-of-the-art results on open-domain QA tasks. However, we argue that current training paradigms harbor a cr...

📖 Read original article


92. Multi-Scale Convolution with Optimal Transport Attention Effect on Multivariate Time Series ​

Author: HaoChong Fu, Jian Xu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10740v1 Announce Type: new Abstract: The analysis of Multivariate Time Series (MTS) plays an important role in a lot of real-world practical applications, but it still remains some challenging problem about capturing multi-granularity structural patterns and suppressing noise appropriatel...

📖 Read original article


93. Policy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT Reasoning ​

Author: Yanmeng Dong, Han Li, Yujia Li, Jingsong Liu, Xun Ma, Yanzhu Hu, Zhengyang Xu, Zhicheng Li, Nassir Navab, Shaohua Kevin Zhou
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10748v1 Announce Type: new Abstract: Computed Tomography (CT) diagnosis often relies on dynamic selection of imaging phases, such as non-contrast, arterial, or venous phases, based on preliminary findings, clinical suspicion, and diagnostic guidelines. This phase-wise decision process is ...

📖 Read original article


94. The VC dimension of partial concept classes via Radon's theorem ​

Author: Grigory Ivanov, Attila Jung, M'arton Nasz'odi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, math.CO, math.FA

arXiv:2607.10751v1 Announce Type: new Abstract: Following Alon, Hanneke, Holzman, and Moran (FOCS 2021), we define a partial concept class (PCC) as a family of partial functions (f: V\to{0,1,\ast}); equivalently, its concepts partition the ground set into black ($f^{-1}(1)$), grey ($f^{-1}(\ast)...

📖 Read original article


95. LSTrans: Efficient Knowledge Transfer for Lightweight and Automated ECG Classification ​

Author: Yi Zhao, Jiajun Gao, Chenyang Xu, Yuxi Zhou, Hao Wang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10784v1 Announce Type: new Abstract: Deploying deep learning models for automated electrocardiogram classification on resource-constrained wearable devices remains challenging due to high computational costs. To address this, we propose LSTrans, a lightweight hybrid model designed for eff...

📖 Read original article


96. Hierarchical Bayesian Quadrature ​

Author: Tim Weiland, Toni Karvonen, Philipp Hennig
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2607.10793v1 Announce Type: new Abstract: Numerical integration is a cornerstone of various scientific computing applications, such as engineering simulations and model evidence computations in probabilistic machine learning. Bayesian Quadrature uses Gaussian process surrogates that explicitly...

📖 Read original article


97. Weight-Adjusted Gradients Reveal Parameter Importance and Failure Modes in LLMs ​

Author: Shrestha Datta, Hongfu Liu, Anshuman Chhabra
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10803v1 Announce Type: new Abstract: Understanding which parameters are influential in Large Language Models (LLMs) is central to improving their efficiency, reliability, and interpretability. We introduce Weight-Adjusted Gradients (WAG), a simple yet effective approach for estimating par...

📖 Read original article


98. When does distribution shift break graph neural networks calibration? ​

Author: Abderaouf Bahi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10804v1 Announce Type: new Abstract: Graph neural networks (GNNs) are increasingly deployed in real-world applications where distribution shift is un-avoidable. However, how such shifts affect model calibration, defined as the agreement between predictive confidence and actual accuracy, r...

📖 Read original article


99. Lower Bound on the Cumulative Constrained Violation for the OGD+Projection algorithm for Constrained Online Convex Optimization (COCO) ​

Author: Haricharan Balasundaram, Karthick Krishna Mahendran, Rahul Vaze
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10808v1 Announce Type: new Abstract: The problem of constrained online convex optimization is considered, where at each round, once a learner commits to an action $x_t \in \mathcal{X} \subset \mathbb{R}^d$, a convex loss function $f_t$ and a convex constraint function $g_t$ that drives th...

📖 Read original article


100. Diachronic Sample Integration: Robust Tail-Risk Estimation with Generative Models ​

Author: Shuning Zhao, Patrick Wong, Leran Zhang, Xiaolin Hu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-fin.RM

arXiv:2607.10810v1 Announce Type: new Abstract: Deep generative models are increasingly used as simulators for downstream decision-making under data scarcity, but in risk-sensitive applications their usefulness depends on rare adverse scenarios rather than typical samples. Standard generative object...

📖 Read original article


101. Graph Neural Networks for RFID-Based Spatial Geometry Inference in Spatial AI Systems ​

Author: Curtis Shull, Merrick Green, Roy Rucker
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10822v1 Announce Type: new Abstract: Indoor spatial understanding remains a fundamental challenge for intelligent systems operating in physical environments. Traditional RFID localization techniques typically estimate positions of tags using signal strength measurements but fail to captur...

📖 Read original article


102. Predictive Divergence Masks for LLM RL ​

Author: Xiangxin Zhou, Jiarui Yao, Penghui Qi, Bowen Ping, Jiaqi Tang, Haonan Wang, Tianyu Pang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10848v1 Announce Type: new Abstract: Reinforcement learning for large language models (LLMs) typically relies on trust-region masks to stabilize off-policy updates. The dominant PPO-style approach uses the sampled-token importance ratio for two criteria: a proximity criterion, which asks ...

📖 Read original article


103. Reliability Scaling Laws for Quantized Large Language Models ​

Author: Sirine Ayadi, S'andor Dar'oczi, Stephan G"unnemann, Bertrand Charpentier
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10855v1 Announce Type: new Abstract: Quantization is a powerful strategy to build capable and resource-efficient large language models (LLMs) by reducing the bitwidth of the parameters. While quantized LLMs achieve state-of-the-art performance on unperturbed inputs using standard predicti...

📖 Read original article


104. Singular perturbations and hierarchical learning in two-layer neural networks ​

Author: C'edric Gerbelot, Jean-Christophe Mourrat
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, math.DS, math.OC

arXiv:2607.10869v1 Announce Type: new Abstract: We study the population gradient flow of an infinitely wide two-layer neural network learning a misspecified single-index model in high dimension. The two layers are optimized jointly, with a perturbative parameter tuning the relative training speed be...

📖 Read original article


105. Infrared Organization and Critical Cognitive Field Formation in Transformer Dynamics ​

Author: Byung Gyu Chae
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10923v1 Announce Type: new Abstract: Large language models exhibit remarkable emergent behaviors, yet the physical mechanism governing their collective dynamics remains poorly understood. Cognitive Field Theory predicts that learning reorganizes the time-scale density of states (TDOS) thr...

📖 Read original article


106. The Spectral Structure of Latent Treatment Effects ​

Author: Hamza Virk, Bijan Mazaheri, Yihren Wu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.10926v1 Announce Type: new Abstract: Identifying heterogeneous treatment effects under unobserved confounding is central in observational causal inference. In proxy models with a discrete latent confounder, prior Synthetic Potential Outcomes (SPO) [Mazaheri-Squires-Uhler '25] recover the ...

📖 Read original article


107. The Singularity Space: A Generative Diffusion Framework for Signal Representation ​

Author: Eli Bar-Yosef, Amir Averbuch, Eli Turkel
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NA, eess.SP, math.NA

arXiv:2607.10930v1 Announce Type: new Abstract: Generative models often represent signals as dense grids of amplitudes, blurring sharp transients that are crucial for the correctness of physical signals. We introduce Singularity Space, a generative framework that represents signals through complex-p...

📖 Read original article


108. Bandit PCA with Minimax Optimal Regret ​

Author: Mo"ise Blanchard, Dmitrii Ostrovskii, Aadirupa Saha
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.10936v1 Announce Type: new Abstract: We study the bandit-feedback version of online principal component analysis (Bandit PCA): in each round $t = 1,\dots,T$, the adversary selects a $d \times d$ symmetric gain matrix $G_t$ with spectrum in $[0,1]$ and rank at most $r$; the learner simulta...

📖 Read original article


109. Sticky Jump Diffusions: A Unifying View of Masked, Continuous, and Hybrid Diffusion ​

Author: Pascal Jutras-Dub'e, Patrick Pynadath, Jeremy Lu, Yuan Gao, Ruqi Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.10951v1 Announce Type: new Abstract: We introduce Sticky Jump Diffusions (SJDs), continuous-time Markov processes on $\mathbb R^d$ whose discrete anchors are token embeddings. In forward time, anchors release their mass at a hazard rate and the released mass diffuses in the continuous amb...

📖 Read original article


110. WSqD: A Horizon-Free Learning Rate Schedule for Large Model Training ​

Author: Jianhao Ma, Yuxin Chen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2607.10959v1 Announce Type: new Abstract: Standard learning rate schedules such as cosine annealing are tied to a fixed training horizon, limiting their ability to accommodate post hoc horizon extension. Warmup-stable-decay (WSD) partially addresses this issue by maintaining a long constant-ra...

📖 Read original article


111. Reinforcement Learning for Execution under Dynamic Fees in a Closed-Loop DEX Simulator ​

Author: Wen-Ting Wang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP, stat.ML

arXiv:2607.10960v1 Announce Type: new Abstract: Trader-facing dynamic fees are increasingly proposed for automated market makers (AMMs), but historical data do not identify how order flow would respond: trader-facing fees do not vary, trader types are latent, and a replayed tape is not a sequential ...

📖 Read original article


112. Efficient Online Proportional Sampling with Applications to Smoothed Online Learning ​

Author: Amirmahdi Mirfakhar, Maria-Florina Balcan, Hedyeh Beyhaghi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CG, cs.GT, econ.TH

arXiv:2607.10963v1 Announce Type: new Abstract: We study the problem of efficient online proportional sampling from a high-dimensional domain under a $\sigma$-smoothed adversary, where the sampling distribution is induced by a dynamically evolving weight function defined over a sequence of piecewise...

📖 Read original article


113. Enhanced Byzantine-Robust Federated Learning Via Truncated-Quadratic Loss for Heterogeneous Data ​

Author: Zhi-Yong Wang, Hao Nan Sheng, Werner Stefan, Hing Cheung So, Linqi Song, Weitao Xu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10970v1 Announce Type: new Abstract: Federated learning distributes data among $n$ clients, making it vulnerable to malicious attacks and data heterogeneity, which together pose challenges for robust learning. To tackle this issue, centered clipping and Huber aggregators have been exploit...

📖 Read original article


114. A Multi-Agent Framework for Zero-Dimensional Reduced-Order Model Planning ​

Author: Bingteng Sun, Hao Yin, Yiling Chen, Renjie Xiao, Lei Xie, Shanyou Wang, Ruonan Wang, Shubao Chen, Qingzong Xu, Lin Lu, Qiang Du, Junqiang Zhu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10994v1 Announce Type: new Abstract: Zero-dimensional reduced-order models (0D ROMs) are central to multi-dimensional design workflows for high-end complex equipment. However, the planning process currently relies on manual expertise, limiting topological exploration and prolonging iterat...

📖 Read original article


115. TabPFN beyond Tabular Data: Calibration and Accuracy on Multimodal Embeddings ​

Author: Jingxiang Zhang, Lujia Zhong, Zijie Zhu, Shuo Huang, Yuang Xu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11007v1 Announce Type: new Abstract: Few-shot multimodal classification commonly attaches a lightweight head, such as $k$-nearest neighbors, logistic regression, or a linear SVM, to a frozen pretrained encoder. Although computationally efficient, these heads can produce poorly calibrated ...

📖 Read original article


116. When the Reward Suite Is Leaky: A Preregistered Causal Contrast of Natural Verifier False Positives in RLVR ​

Author: Chuyifei Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.11022v1 Announce Type: new Abstract: The test suites used as RLVR rewards for code have natural false positives: per-task, persistent, asymmetric errors that accept the same wrong programs every time they appear, unlike the symmetric or resampled noise assumed by existing noise-robustness...

📖 Read original article


117. Domain-Aware Scaling Laws Uncover Data Synergy ​

Author: Kimia Hamidieh, Lester Mackey, David Alvarez-Melis
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.11052v1 Announce Type: new Abstract: Machine learning progress is often attributed to scaling model size and dataset volume, yet the composition of data can be just as consequential. Empirical findings repeatedly show that combining datasets from different domains yields nontrivial intera...

📖 Read original article


118. AeroMELD: A Linear Embedding of Aerosol Populations for Diagnostics and Latent Dynamics ​

Author: Ehsan Saleh, Saba Ghaffari, Wenhan Tang, Jeffrey H. Curtis, Lekha Patel, Peter A. Bosler, Nicole Riemer, Matthew West
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2607.11073v1 Announce Type: new Abstract: Accurately representing atmospheric aerosol populations is essential for simulating aerosol-cloud interactions, radiative forcing, and ice nucleation, yet existing reduced schemes impose structural assumptions that limit their ability to capture compos...

📖 Read original article


Author: Vignatha Vinjam, Manjunath Kolavennu, Myna Vajha, Karthik Periyapattana Narayanaprasad
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11075v1 Announce Type: new Abstract: The choice of Modulation and Coding (MCS) type for a particular channel condition is made through link adaptation (LA) algorithms that operate at the MAC layer. These algorithms rely on the ACK/NACK statistics and the channel quality index (CQI) feedba...

📖 Read original article


120. Adapting Evidential Neural Networks to Test-Time Neighbor Fusion Improves Molecular Property Prediction ​

Author: Cameron Gruich, Weichi Yao, Yixin Wang, Bryan Goldsmith
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM, stat.ML

arXiv:2607.11091v1 Announce Type: new Abstract: A trained molecular property model can be refined at test time by correcting each prediction with the measured labels of the most similar training molecules, a retraining-free procedure we call neighbor fusion; evidential neural networks make it princi...

📖 Read original article


121. Multi-dimensional training-priority weighting based on physical information propagation paths: a unified residual-weighting framework for physics-informed neural networks ​

Author: Zhangyi Lian, Xinda Dong, Wenxuan Huo, Weifeng Huang, Gang Zhu, Qiang He
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11094v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have shown promise for solving partial differential equations (PDEs); however, their synchronous optimization treats residuals of different regions and constraints equally, which is inconsistent with the progres...

📖 Read original article


122. A Novel Graph Fraud Detector via Grouped Attribute Completion and Confidence-Aware Contrastive Learning ​

Author: Junpeng Wu, Ye Yuan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11107v1 Announce Type: new Abstract: Graph fraud detection plays a pivotal role in safeguarding the security and integrity of modern digital ecosystems. Graph Neural Networks (GNNs) are commonly adopted for graph fraud detection. However, the practical performance of existing GNN-based de...

📖 Read original article


123. Neural Discovery of Memory and Nonlocal Kernels in Integro-Differential Equations with Constrained Kolmogorov--Arnold Networks ​

Author: Aruzhan Tleubek, Salah A Faroughi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2607.11110v1 Announce Type: new Abstract: Discovering the memory or nonlocal kernel governing an integro-differential equation (IDE) from sparse and noisy observations is an ill-posed inverse problem. Existing identification methods often rely on problem-specific analytical derivations, specia...

📖 Read original article


124. CA-DGCL: Dynamic Graph Continual Learning via Condensation and Attachment ​

Author: Tingxu Yan Ye Yuan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11112v1 Announce Type: new Abstract: Dynamic graph continual learning (DGCL) is an effective manner for handling catastrophic forgetting in dynamic graphs. However, existing DGCL methods underutilize temporal information across graph snapshots. To address this critical issue, we propose a...

📖 Read original article


125. The Equilibrium Is the Initialization: Lazy Identity Collapse in Physics-Structured Deep Equilibrium Reasoning ​

Author: Joyjeet Singh
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11116v1 Announce Type: new Abstract: Deep equilibrium models promise input-adaptive implicit computation: harder problems should demand more solver iterations, and the solved equilibrium should encode the result of genuine iterative inference. We report a cautionary study of a port-Hamilt...

📖 Read original article


126. ToolAtlas: Learning Once, Reusing Everywhere with Tool-Side Memory ​

Author: Yue Fang, Zhibang Yang, Fangkai Yang, Xiaoting Qin, Liqun Li, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11126v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external tools served by shared providers and accessed by heterogeneous downstream agents. Existing approaches improve tool use on the agent side through parameter updates, prompt refinement, or ag...

📖 Read original article


127. Learning Subgroup Relations Using Siamese Graph Neural Networks ​

Author: Tal Weissblat
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, math.GR

arXiv:2607.11140v1 Announce Type: new Abstract: Determining whether one finite group is isomorphic to a subgroup of another is a fundamental problem in computational group theory. In this work, we propose a Siamese Graph Neural Network (Siamese GNN) for subgroup prediction using Cayley graph represe...

📖 Read original article


128. Rank-Conditioned Sample Reuse for the Plackett--Luce Best-of-$K$ Objective ​

Author: Melveena Jolly, Midhun Xavier
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2607.11146v1 Announce Type: new Abstract: We study the coupled objective J_K^WOR = E_{S ~ PL-WOR_K}[max_{i in S} R_i]: the expected maximum reward of a size-K Plackett-Luce draw without replacement, the law of Gumbel-Top-K / Stochastic Beam Search decoding. This estimand differs from the conve...

📖 Read original article


129. NeuroMem-FHP: A Likelihood-Free Deep Learning Framework for Parameter Estimation of Fractional Hawkes Process ​

Author: Neha Gupta, Aditya Maheshwari
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.11177v1 Announce Type: new Abstract: In this paper, we propose deep learning based NeuroMem-FHP framework for estimating the parameters of the fractional Hawkes process (FHP), a self-exciting point process that captures long-range dependence through a fractional Mittag-Leffler excitation ...

📖 Read original article


130. FastTPS: An Optimized Method for LLM Token Phase for AI accelerators ​

Author: Wenzong Yang, Danyang Zhang, Kun Cao, Tejus Siddagangaiah, Rajeev Patwari, Zhanxing Pu, Siyin Kong, Zijiang Yang, Hao Zhu, Varun Sharma, Yue Gao, Tianping Li, Fan Yang, Jicheng Chen, Yushan Chen, Fennian Zhao, Aaron Ng, Elliott Delaye, Ashish Sirasao, Sudip Nag
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11211v1 Announce Type: new Abstract: The popularity of large language models (LLMs) escalates an ongoing demand for effective inference. However, due to the sequential processing of tokens during the token phase in decoder-only LLMs inference, the inherent low parallelism leads to reduced...

📖 Read original article


131. Trustworthy synthetic data for campaign decision support: strategy simulation fidelity and the PolicySynth framework ​

Author: Tung Dang, The Hung Phung, Son Lam Nguyen, Tu Nguyen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.11269v1 Announce Type: new Abstract: Decision support systems (DSS) increasingly run retention what-if analysis on synthetic customer populations, because privacy constraints preclude unrestricted use of real data. Such a system is trustworthy only if the synthetic data lead managers to t...

📖 Read original article


132. SPARC-Net: A Spectral, Causality-Aware, and Hard-Constrained Physics-Informed Architecture for Stiff and Shock-Dominated Partial Differential Equations ​

Author: Divyavardhan Singh, Dimple Sonone, Hammad Mohammad, Kishor Upla
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.DM

arXiv:2607.11310v1 Announce Type: new Abstract: Physics-Informed Neural Networks (PINNs) provide a meshless approach for solving partial differential equations (PDEs), but suffer severe degradation in stiff and shock-dominated problems, where small PDE residuals can correspond to globally inaccurate...

📖 Read original article


133. PRISM Edit: One Vector for All Temporal Answers ​

Author: Chen Huang, Qi Zheng, Ruiqin Zheng, Long Zeng, Yuantong Xu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11327v2 Announce Type: new Abstract: Model editing keeps large language models (LLMs) up to date without retraining, but temporal facts expose a limitation of the prevailing locate-and-edit paradigm: an update is not always a replacement. When a fact changes, the new answer should become ...

📖 Read original article


134. BackgroundMellow: A Multi-Modal Cohesive Framework for Narrative-Driven Rich Cinematic Soundscape Generation ​

Author: Ajitesh Jamulkar, Aritra Hazra
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MM

arXiv:2607.11364v1 Announce Type: new Abstract: Generating immersive, synchronized and cinematic audio for long-form textual narratives remains a significant challenge in multi-modal AI. While current Text-to-Audio (TTA) frameworks successfully synthesize isolated sound effects, they struggle with n...

📖 Read original article


135. Surprisingly Simple and Effective Multi-Domain Graph Foundation Model through Graph-to-Table Alignment ​

Author: Chunyu Hu, Tianyin Liao, Ge Lan, Xingxuan Zhang, Jianxin Li, Peng Cui, Ziwei Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11374v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) have emerged as a promising paradigm for learning transferable representations across diverse graph domains. Recent advancements in GFMs have been largely dominated by two paradigms: Graph Neural Network and Large Languag...

📖 Read original article


136. Physics-Aware Conditional SetGAN for Spatially Consistent Multi-User TR 38.901 Channel Generation ​

Author: Mauro Gonzalo Tarazona-Levano, David Lopez-Perez, Nicola Piovesan, David Gomez-Barquero
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11429v1 Announce Type: new Abstract: TR 38.901-based channel models such as Sionna are reliable, but generating many multi-user channel realizations remains expensive. This paper asks a practical question: can a trained generative model produce multi-user TR 38.901 channels faster than Si...

📖 Read original article


137. Generalizing Preference-based Reinforcement Learning: a Rationality Model for Incomparability ​

Author: Simone Drago, Marco Mussi, Leonardo Bianconi, Alberto Maria Metelli
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11432v1 Announce Type: new Abstract: In this work, we study the reinforcement learning (RL) problem from pairwise trajectory comparisons provided by a human expert. We generalize preference-based RL by formalizing a novel setting in which the expert can also label trajectory pairs as inco...

📖 Read original article


138. Velocity Scheduled Flow Matching ​

Author: Vitalii Bondar
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11442v1 Announce Type: new Abstract: Flow matching trains a neural network to regress the conditional velocity along a linear interpolant between noise and data, and the number of network evaluations~(NFE) sets the cost of sampling. The straight-line interpolant carries an implicit choice...

📖 Read original article


139. Event-based Neural Decoding for Neuroprosthetic Motor Control ​

Author: Khaleelulla Khan Nazeer, Sirine Arfa, Matthias Jobst, Richard George, Christian Mayr
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2607.11445v1 Announce Type: new Abstract: A substantial number of patients experience diminished mobility due to disabilities, diseases, or accidents. Although modern prostheses, powered by deep neural networks, hold the promise of significantly enhancing the quality of life for these individu...

📖 Read original article


140. HyperSafe: Inference-Time Safety Recovery for Fine-Tuned Language Models ​

Author: Aznaur Aliev, Carlos Hinojosa, Abdelrahman Eldesokey, Bang An, Bernard Ghanem, Yibo Yang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.11475v1 Announce Type: new Abstract: Safety alignment in large language models can be fragile under fine-tuning, as even benign task adaptation may increase harmful compliance. Existing defenses mainly follow two directions: they either intervene during or after fine-tuning through retrai...

📖 Read original article


141. Agentic Skill Optimization over Lie Algebroids ​

Author: Sridhar Mahadevan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.CT

arXiv:2607.11493v1 Announce Type: new Abstract: Agentic systems increasingly improve themselves by editing skills: prompts, rubrics, plans, tool contracts, examples, validators, and traces. Skill edits are not independent coordinates in a vector space: they are local repairs to structured artifacts ...

📖 Read original article


142. IG-GAN: A Generative Adversarial Network for Aerodynamic Data Generation Based on Intrinsic Geometry ​

Author: Ying Yan, Liwei Hu, Xiaoming Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11497v1 Announce Type: new Abstract: Existing generative models learn data distributions in flat Euclidean space. However, most data in our real world are manifolds embedded in high dimensional Euclidean space. Therefore, we propose an intrinsic-geometry-based generative adversarial netwo...

📖 Read original article


143. Proxy Exploration and Reusable Guidance: A Modular LLM Post-Training Paradigm via Proxy-Guided Update Signals ​

Author: Daocheng Fu, Rong Wu, Yu Yang, Xuemeng Yang, Jianbiao Mei, Licheng Wen, Pinlong Cai, Yong Liu, Botian Shi, Yu Qiao
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11505v1 Announce Type: new Abstract: Post-training is essential for refining the domain-specific capabilities of large language models (LLMs), yet existing reward optimization and distribution matching methods tightly couple policy exploration with distribution alignment. This coupling fo...

📖 Read original article


144. SCOPE-RL: Optimizing Reasoning Paths Before and After Success ​

Author: Xiaojian Liu, Han Xu, Jianqiang Xia, Zhixuan Li, Ke Xu, Yiwei Dai, Xinran Chen, Changwo Wu, Yuchen Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.11506v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) optimizes LLMs using sparse verifiable final-answer rewards. This sparse anchor reliably verifies whether a trajectory succeeds but provides no direct feedback on the reasoning path that produced it...

📖 Read original article


145. CDFM: Towards a General-Purpose Causal Discovery Foundation Model ​

Author: Jie Qiao, Ruichu Cai, Zijian Li, Weilin Chen, Pengfei Hua, Boyan Xu, Zhengming Chen, Zhifeng Hao, Peng Cui
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.11508v1 Announce Type: new Abstract: Causal discovery, the process of recovering underlying causal structures from observational data, is a fundamental pursuit across scientific disciplines. Over the past decades, numerous algorithms have been developed to tackle this challenge through wo...

📖 Read original article


146. DAG-FM: A Foundation Model for Causal Discovery under Heterogeneous Causal Mechanisms ​

Author: Yikang Chen, Zhengkang Guan, Haoyuan Qian, Peng Cui, Yi Yang, Kun Kuang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.11510v1 Announce Type: new Abstract: Causal discovery from observational tabular data remains fundamentally challenging, primarily due to the heterogeneity of underlying causal mechanisms and the high-dimensional combinatorial search space of Directed Acyclic Graphs (DAGs). In this paper,...

📖 Read original article


147. AutoMatBench: An Automatic Optimization Toolkit for the Acceleration of Material Properties Prediction Benchmarking ​

Author: Hongxiao Li, Wanling Gao
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11526v1 Announce Type: new Abstract: Material property prediction (MPP) infers key properties from chemical composition and structure, accelerating the discovery and optimization of novel materials. In the realm of MPP, MatBench is a widely accepted benchmarking tool that defines over ten...

📖 Read original article


148. Random Label Prediction Heads for Studying Memorization in Deep Neural Networks ​

Author: Marlon Becker, Jonas Konrad, Luis Garcia Rodriguez, Benjamin Risse
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11541v1 Announce Type: new Abstract: We introduce a straightforward yet effective method to empirically study memorization in deep neural networks for classification tasks. Our approach augments each training sample with auxiliary random labels, which are then predicted by a random label ...

📖 Read original article


149. Condition-Stratified Robustness Analysis of Post-Hoc Calibration Methods for Probabilistic Classifiers ​

Author: Gurdeep Singh Virdee
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11542v1 Announce Type: new Abstract: Post-hoc calibration is widely adopted to correct probability estimates from trained classifiers, yet most evaluations report aggregate performance without testing whether that performance holds across distinct operating conditions within a single data...

📖 Read original article


150. Advancing Optimal Subset Oracle via Learning Relaxation of Neural Set Functions ​

Author: Yongquan Shi, Zijing Ou, Shiping Wang, Yatao Bian
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11555v1 Announce Type: new Abstract: Learning neural set functions is pivotal to a wide range of important applications, including compound selection in AI-driven drug discovery and product recommendation. Recent work has introduced optimal subset oracles to implicitly learn set functions...

📖 Read original article


151. Heuristic Learning for Active Flow Control Using Coding Agents ​

Author: Paul Garnier, Jonathan Viquerat, Elie Hachem
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, physics.flu-dyn

arXiv:2607.11565v1 Announce Type: new Abstract: Active flow control involves nonlinear dynamics, partial observations, and computationally expensive simulations, making controller design particularly challenging. Deep reinforcement learning (DRL) has emerged as a powerful framework for such problems...

📖 Read original article


152. Structure-Feature Aligned Graph Learning via Alternating Constrained Optimization ​

Author: Chengcheng Yan, Qingsong Wang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11577v1 Announce Type: new Abstract: We introduce a constrained two-view framework for node prediction that aligns structure-conditioned GNN embeddings with a structure-free feature prior learned by an anchor model. Conventional Graph Neural Networks (GNNs) couple feature transformation a...

📖 Read original article


153. DiffEEG: A Self-Supervised Denoising Diffusion Model for Learning EEG Generic Representations ​

Author: Abdulkader Helwan, Lina Abou-Abbas, Hussein El Amouri, Belkacem Chikhaoui, Khadidja Henni
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, eess.SP

arXiv:2607.11578v1 Announce Type: new Abstract: Deep learning for EEG-based seizure detection faces critical challenges: severe annotation scarcity and extreme class imbalance, where ictal events comprise less than 10% of clinical recordings. We present DiffEEG, a 9.6M-parameter self-supervised fou...

📖 Read original article


154. Privacy-Aware Collaborative and Distributed Bayesian Optimization ​

Author: Aditya Rane, Sathwik Yamana, Paritosh Ramanan, Srikanthan Ramesh, Akash Deep
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ME

arXiv:2607.11600v1 Announce Type: new Abstract: We propose a collaborative meta-learning framework for distributed Bayesian optimization matching centralized performance without raw-data exchange. We show gradient sharing leaks client observations, with leakage worsening as the search converges and ...

📖 Read original article


155. Fundamental Limitations of Fixed-Budget Best-Arm Identification ​

Author: Motti Goldberger
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.11635v1 Announce Type: new Abstract: In fixed-budget best-arm identification, also known as ranking and selection, an algorithm has a sampling budget to distribute across $K$ arms. Each sample provides noisy feedback about that arm's mean, and the goal is to identify the arm with the larg...

📖 Read original article


156. Bet on Features: Anytime-Valid and Feature-Aware Auditing of Conditional Quantile Forecasters ​

Author: Ivane Antonov, Sohom Mukherjee, Richard Pibernik, Yo Joong Choe
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.11653v1 Announce Type: new Abstract: Black-box conditional quantile forecasts are widely used for sequential decisions under asymmetric costs, such as inventory planning in supply chain management. Once deployed, such forecasters must be monitored continuously as data streams drift and re...

📖 Read original article


157. How to Tame Grokking: Representation Geometry as a Control Signal ​

Author: Maksim A Kazanskii
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11666v1 Announce Type: new Abstract: Grokking is a phenomenon in which neural networks initially memorize training data and only later exhibit strong generalization after prolonged optimization. Despite extensive recent study, the factors influencing the emergence and timing of grokking r...

📖 Read original article


158. A multi-scale feature enhanced graph neural network for fluid dynamics prediction in complex geometries ​

Author: Li Xiao, Tianyu Li, Yiye Zou, Mingjie Zhang, Xiaogangd Deng
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn

arXiv:2607.11672v1 Announce Type: new Abstract: Industrial design in fields such as vehicle and aerospace engineering often relies on large-scale numerical simulations to evaluate fluid dynamics performance, which can incur substantial computational costs. Deep neural networks have shown promise in ...

📖 Read original article


159. CatRetriever: Contrastive Representation Learning for Slab-to-Bulk Retrieval in Generative Catalyst Discovery ​

Author: Jungho Oh, Woosung Kim, Dong Hyeon Mok, Jonggeol Na, Seoin Back
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci

arXiv:2607.11712v1 Announce Type: new Abstract: Inverse design is an emerging data-driven paradigm for efficiently navigating vast chemical spaces to discover new materials with targeted properties, and in the context of heterogeneous catalysis, surface generative models have recently advanced this ...

📖 Read original article


160. Active Offline-to-Online Reinforcement Learning ​

Author: Alper Kamil Bozkurt, Shangtong Zhang, Yuichi Motai
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11720v1 Announce Type: new Abstract: Background: Offline reinforcement learning (RL) enables effective policies to be trained from large, previously collected datasets and subsequently improved through limited online interaction. This offline-to-online RL (O2O-RL) paradigm is particularly...

📖 Read original article


161. Time-Lag-Aware Deep Reinforcement Learning for Flexible Job-Shop Scheduling in PPVC Module Factories ​

Author: Ziheng Zhang, Wei Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC

arXiv:2607.11725v1 Announce Type: new Abstract: Prefabricated prefinished volumetric construction moves most building work into module factories, whose production floor operates as a flexible job shop. A major complication is decisive: long post-operation time-lags caused by concrete curing, waterti...

📖 Read original article


162. HiFi-LLP: High-Fidelity, Low-Cost Latency Predictors with Confidence for Robust HW-NAS ​

Author: Shambhavi Balamuthu Sampath, Behzad Shomali, Nael Fasfous, Moritz Thoma, Judeson Anthony Fernando, Lukas Frickenstein, Pierpaolo Mori, Manoj Rohit Vemparala, Alexander Frickenstein, Walter Stechele
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AR

arXiv:2607.11746v1 Announce Type: new Abstract: With deep neural networks (DNNs) increasingly deployed on edge devices, hardware (HW)-aware optimization techniques--such as HW-aware compression and HW-aware neural architecture search (HW-NAS)--have become essential. These methods rely on real feedba...

📖 Read original article


163. From Global to Factor-Wise Expert Composition in Discrete Diffusion Models ​

Author: Haozhe Huang, Yudong Xu, Abhijoy Mandal, Al'an Aspuru-Guzik
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11758v1 Announce Type: new Abstract: Discrete diffusion models offer a powerful framework for solving complex reasoning tasks, particularly through compositional generation, which combines multiple pre-trained experts to generalize beyond their individual training data. Recent theoretical...

📖 Read original article


164. From Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASP ​

Author: Michael Rizvi-Martel, Satwik Bhattamishra, Guillaume Rabusseau, Michael Hahn
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.11760v1 Announce Type: new Abstract: A theoretical understanding of Transformers is crucial to better understand the capacities and limitations of large language models (LLMs). There is much work analyzing the expressivity of attention-based models. By proposing handcrafted weights or usi...

📖 Read original article


165. An Exact Instrument for State Usage in Selective State-Space Models, and the Input-Driven Migration It Reveals ​

Author: Raktim Bhattacharya
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11796v1 Announce Type: new Abstract: Selective state-space models such as Mamba route information through a bank of first-order modes whose input coupling is set by a learned selection mechanism. We give an exact instrument for measuring how a trained model uses these modes. Because the s...

📖 Read original article


166. Relaxing Faithfulness with Intervention-Only Causal Discovery ​

Author: Bijan Mazaheri, Jiaqi Zhang, Caroline Uhler
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.11816v1 Announce Type: new Abstract: Causal discovery algorithms learn a network that describes the causal dependencies among random variables. A common workflow involves first utilizing conditional independence properties on observational data to determine partially directed causal relat...

📖 Read original article


Author: Romain Amigon
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE

arXiv:2607.11826v1 Announce Type: new Abstract: Neural Architecture Search (NAS) has automated the design of deep learning models but traditionally requires massive computational resources, often measured in thousands of GPU-days. In this paper, we propose a frugal and memetic NAS framework designed...

📖 Read original article


168. Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias ​

Author: Zixiang Xu, Sixian Li, Huaxing Liu, Xiang Wang, Shuai Li, Zirui Song, Xiuying Chen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.11871v1 Announce Type: new Abstract: Existing studies of LLM-as-judge scoring bias work predominantly at the input-output level: they perturb inputs, measure score deltas, and propose prompt-level mitigations. We argue that the same biases admit a representation-level account in the judge...

📖 Read original article


169. Invariant Learning Dynamics of Transformers in Inductive Reasoning Tasks ​

Author: Tiberiu Musat, Tiago Pimentel, Nicolas Zucchet, Thomas Hofmann
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11875v2 Announce Type: new Abstract: We present a theoretical framework to explain the emergence of inductive reasoning abilities in Transformer language models. While previous works on Transformer learning dynamics have so far been mostly tied to specific tasks, we study a generalized cl...

📖 Read original article


170. Requential Coding: Pushing the Limits of Model Compression with Self-Generated Training Data ​

Author: Shikai Qiu, Marc Finzi, Yujia Zheng, Kun Zhang, Andrew Gordon Wilson
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11883v1 Announce Type: new Abstract: Compression is fundamental to intelligence. A model that can represent its training data as a short code has discovered regularities that enable generalization. Large neural networks may learn functions far simpler than their parameter counts suggest, ...

📖 Read original article


171. Format Sensitivity Index: Token-Controlled Prompt Wrapper Robustness and Schema Compliance in LLM Benchmarking ​

Author: Deep Pankajbhai Mehta
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.09665v1 Announce Type: cross Abstract: Prompt wrappers often differ only in formatting, yet they can change model scores enough to flip leaderboard conclusions. We study this variance under a token-controlled protocol and introduce two complementary metrics: the Format Sensitivity Index (...

📖 Read original article


172. Faithful, Not Corrective: Message-Format Effects in Multi-Hop Agent Relays Are Tier-Dependent ​

Author: Zayx Shawn
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.09678v1 Announce Type: cross Abstract: When LLM agents hand off information to one another, does the message format matter? Two literatures disagree: format-optimization work reports that structured messages cut cost without hurting accuracy, while format-restriction work finds that impos...

📖 Read original article


173. ECG-LDC: A Hardware-Efficient Low-Dimensional Computing Framework for ECG Arrhythmia Classification ​

Author: Anh Tran, Khanh Tran, Cuong Do
Published: 7/14/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG

arXiv:2607.09680v1 Announce Type: cross Abstract: Continuous cardiac monitoring in wearable devices demands classifiers that are simultaneously accurate, energy-efficient, and deployable on resource-constrained hardware. While deep neural network approaches have demonstrated high classification accu...

📖 Read original article


174. Transfer Learning Across Policy Regimes in Adaptive Multi-Agent Systems ​

Author: Roberto Garrone
Published: 7/14/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.CY, cs.LG

arXiv:2607.09685v1 Announce Type: cross Abstract: Policy models often assume that the relationship between a policy instrument and its outcome remains stable across institutional conditions. In adaptive socio-technical systems this assumption may fail: regulatory change can alter incentives, agents ...

📖 Read original article


175. Interpreting Latent CoT Reasoning as Dynamical Systems ​

Author: Sabari Iyyappan Duraipandian, Shreya Sanjay Boyane, Manju Nagesh, Jerome Francis, Archana Vaidheeswaran, Kevin Zhu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.09698v1 Announce Type: cross Abstract: Recent latent reasoning methods, such as CODI and COCONUT, face a fundamental interpretability problem: they maintain multiple superimposed candidate traces in the hidden space at each step, unlike explicit- CoT, which follows a single transparent re...

📖 Read original article


176. Model Collapse: On Recursion, Noise, and Uncharted Machine Visions ​

Author: Violaine Boutet de Monvel (LIRA, IRCAV)
Published: 7/14/2026, 4:00:00 AM
Categories: cs.OH, cs.AI, cs.CY, cs.LG

arXiv:2607.09705v1 Announce Type: cross Abstract: Since 2023, computer scientists have warned against model collapse -- the contamination of training sets with AI-generated outputs that progressively degrade model performance. Exemplifying a positive-feedback-driven failure, it produces effects such...

📖 Read original article


177. YUKTI: From Natural-Language Situations to Robust, Verifiable Decisions An Uncertainty-Typed Proposition IR, Assumption-Robust Pareto Frontiers, and a Regret Certificate ​

Author: Suyash Mishra
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.GT, cs.LG

arXiv:2607.09706v1 Announce Type: cross Abstract: Language models turn a worded situation into a numeric plan, and the dominant pipelines (NL4Opt, OptiMUS, ORLM, OR-LLM-Agent) commit to a single objective and point-valued coefficients, then solve once. For decisions that allocate real budget, effort...

📖 Read original article


178. GES-TSP: Graph Edge Sparsification for TSP ​

Author: Tianfeng Chen, Xianyue Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, math.CO

arXiv:2607.09708v1 Announce Type: cross Abstract: Solving large-scale instances of the Traveling Salesman Problem (TSP) exactly is computationally expensive. Researchers often employ graph sparsification methods to improve computational efficiency. Traditional sparsification methods typically rely o...

📖 Read original article


179. Saturation-Aware Robust Trajectory Optimization for Reusable Launch Vehicles via Differentiable Physics ​

Author: Liwei Chen, Tong Qin
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, physics.app-ph

arXiv:2607.09736v1 Announce Type: cross Abstract: The high-angle-of-attack flip maneuver of reusable launch vehicles presents significant challenges for robust trajectory optimization due to the combined effects of highly nonlinear dynamics, aerodynamic uncertainties, and actuator saturation. This p...

📖 Read original article


180. Q-Score: A Quantum-Native Scoring Function for Molecular Docking ​

Author: Kangyu Zheng, Yidong Zhou, Ruihao Li, Zixin Ding, Zhiding Liang, Shaohua Li
Published: 7/14/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG

arXiv:2607.09737v1 Announce Type: cross Abstract: Molecular docking predicts how a small molecule binds to a protein and is a key bottleneck in drug discovery. Classical scoring functions sum empirical pairwise contacts, blind to quantum-mechanical effects like orbital charge transfer that govern bi...

📖 Read original article


181. SupplyNetPy: An Open-Source Python Library for High-Fidelity Modeling and Simulation of Arbitrary Supply Chain and Inventory Networks ​

Author: Tushar Lone, Neha Karanjkar
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.09745v1 Announce Type: cross Abstract: This paper introduces SupplyNetPy, an open-source, well-documented Python library for modeling and discrete-event simulation of supply chain networks with arbitrary multi-echelon structures. It supports multiple replenishment policies, perishable inv...

📖 Read original article


182. MorphologyFM: A Foundation Model for Morphology-Aware Representation Learning from ECG and Pulse Oximetry Waveforms ​

Author: Saiyang Feng, Yuanyun Zhang, Shi Li
Published: 7/14/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG

arXiv:2607.09749v1 Announce Type: cross Abstract: Foundation models have recently emerged as a powerful paradigm for learning transferable representations from large scale biomedical data, yet existing approaches for physiological waveforms primarily optimize reconstruction or forecasting objectives...

📖 Read original article


183. Task-Conditioned Synthetic Data Generation for Improving Machine Learning Performance in Agricultural Prediction Tasks ​

Author: Hamid Ebrahimy, Moritz Lucas, Martin Atzmueller
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.09751v1 Announce Type: cross Abstract: Machine Learning (ML) algorithms have been widely used to estimate agricultural variables across diverse contexts. However, because the quantity and quality of training data strongly influence performance of ML algorithms, their use can be constraine...

📖 Read original article


184. Unified Backbone Refinement for Diffusion Models via Internal-Latent Analysis ​

Author: Haksoo Lim, Myeongjin Lee, Wonjoon Chang, Jaesik Choi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.09753v1 Announce Type: cross Abstract: Diffusion models have achieved remarkable success across diverse domains, with performance closely related to the denoising backbones that parameterize the score function. In this paper, we present a systematic, phase-aware analysis of diffusion comp...

📖 Read original article


185. Cross-Subject Modeling for Widefield Calcium Imaging via Atlas-Aligned Spatiotemporal Tokenization ​

Author: Mohammad Hosseini, Eray Erturk, Saba Hashemi, Maryam M. Shanechi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, q-bio.NC

arXiv:2607.09754v1 Announce Type: cross Abstract: Large-scale, multi-subject widefield calcium imaging provides unprecedented access to brain-wide cortical dynamics. However, the high dimensionality, complex spatiotemporal structure, and substantial task-irrelevant activity in widefield recordings h...

📖 Read original article


186. Knowledge-Constrained Shape Optimization with a Mixture-of-Experts Neural Operator for High-Confidence Design ​

Author: Wenhao Fan, Yuanwei Bin, Jianghan Gu, Wenfa Luo, Jiao Xiang, Yuntian Chen, Shiyi Chen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, physics.comp-ph

arXiv:2607.09763v1 Announce Type: cross Abstract: Engineering shape optimization faces challenges in both expert-dependent problem setup and surrogate-model reliability. In practical aerodynamic design, optimization settings such as editable regions, deformation ranges, and design-preservation const...

📖 Read original article


187. Norm Enforcement for AI Agents: Robustly Shaping Behavior in Multi-Agent Systems ​

Author: Yaowen Ye, Jacob Steinhardt
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2607.09766v1 Announce Type: cross Abstract: AI agents are increasingly deployed in shared environments where they pursue diverse goals and compete for rewards. This multi-agent competition can lead to behaviors that serve individual gains at collective cost -- for instance, marketing agents ma...

📖 Read original article


188. A Risk-Field Enhanced Closed-Loop Digital Twin Framework for Autonomous Driving Safety Validation ​

Author: Yongzhi Liu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.09772v1 Announce Type: cross Abstract: Autonomous driving systems require reliable safety validation before real-world deployment. However, large-scale road testing is costly, difffcult to reproduce, and inefffcient for exposing rare safety-critical scenarios. Conventional simulation impr...

📖 Read original article


189. EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents ​

Author: Mianqiu Huang, Taofeng Xue, Chong Peng, Jinrui Ding, Sicheng Fan, Jiale Hong, Yufei Gao, Xiaocheng Zhang, Linsen Guo, Xin Yang, Dengchang Zhao, Yuchen Xie, Peng Pei, Xunliang Xie, Xipeng Qiu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.09773v1 Announce Type: cross Abstract: Computer-use agents must solve long-horizon tasks through repeated interaction with partially observable, multimodal desktop environments. Although imitation learning and offline trajectory refinement provide strong priors, static traces cannot cover...

📖 Read original article


190. Towards Real-World Wearable Motion Reconstruction ​

Author: Andrea Boscolo Camiletto, Rishabh Dabral, Eduardo Alvarado, Thabo Beeler, Marc Habermann, Christian Theobalt
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.09780v1 Announce Type: cross Abstract: The modern-day surge in popularity of wearable devices poses a fundamentally unique motion capture problem: reconstructing full-body movement from any set of sensing hardware worn at a given moment. Yet, most research efforts assume fixed sensor conf...

📖 Read original article


191. Length Penalties Make Chain-of-Thought Less Monitorable ​

Author: Bryce Little
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.09786v1 Announce Type: cross Abstract: Length-penalized reinforcement learning can shorten chain-of-thought reasoning while hiding an influence that drives the model's answer. In our experiments, training with length penalties does not stop misleading hints from steering models, even thou...

📖 Read original article


192. Adversarially Guided Diffusion for LiDAR Range Image Synthesis ​

Author: Stavros Bouras, Antonios Makris, Alexandros Gkillas, Aris S. Lalos, Konstantinos Tserpes
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.09787v1 Announce Type: cross Abstract: LiDAR semantic segmentation is a key perception task in autonomous driving, where false predictions can affect downstream planning and safety-critical decision-making. Although adversarial attacks, and specifically adversarial examples, have been wid...

📖 Read original article


193. MVMGNN;Multi-View Masked Graph Neural Network for Alzheimer's Disease Diagnosis using Structural MRI ​

Author: Ni Yao, Zhenxu Wang, Danyang Sun, Chuang Han, Yanting Li, Jiaofen Nan, Fubao Zhu, Chen Zhao, Weihua Zhou
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.09788v1 Announce Type: cross Abstract: Alzheimer's disease (AD) is a common neurodegenerative disorder, and early diagnosis is of great significance for delaying disease progression and enabling timely intervention. Mild cognitive impairment (MCI), which represents an intermediate clinica...

📖 Read original article


194. CHM-Net: Center Heatmap-driven Macro-Micro Modeling Network for MRI-based Microbial Density Stratification ​

Author: Jiaming Liang, Haolin Chen, Tingting Li, Bowen Yu, Qianyan Long, Tinghe Zhang, Xi Zhong, Xiaowei Hu, Xiaoqi Sheng, Hongmin Cai
Published: 7/14/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2607.09812v1 Announce Type: cross Abstract: Microbial density is clinically important for tumor assessment and treatment decision-making, and recent advances in deep learning suggest that it can be non-invasively inferred from multimodal MRI. In this work, MRI-based Microbial Density Stratific...

📖 Read original article


195. Generative Testing of Automated Speech Recognition Systems ​

Author: Yanis Xabier Wilbrand Pe~na, Oliver Wei{\ss}l, Andrea Stocco
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2607.09833v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) systems have achieved high accuracy with transformer-based models, enabling deployment in critical applications. However, they remain vulnerable to adversarial manipulation, particularly in black-box settings where ...

📖 Read original article


196. ShapKO: Shapley-Adaptive Modality Knockout for Robust Multimodal Learning ​

Author: Nusrat Binta Nizam, Fengbei Liu, Sunwoo Kwak, Minh Nguyen, Ruining Deng, Mert R. Sabuncu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.09884v1 Announce Type: cross Abstract: Multimodal medical models often degrade when inputs are missing, a common scenario in real-world clinical workflows. Separately, even when all modalities are present, modality dominance is observed during training, where optimization over-relies on a...

📖 Read original article


197. An End-to-End Hybrid Quantum--Classical Sampling Workflow for Discrete Markov Random Fields: A Reproducible Case Study ​

Author: Arul Rhik Mazumder
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.09893v1 Announce Type: cross Abstract: Sampling from discrete Markov random fields (MRFs) is a hard problem. We study amplitude-encoded i.i.d. sampling for small MRFs where $2^n$ target probabilities are precomputed classically. This removes quantum exponential speedup but allows a clean ...

📖 Read original article


198. When Classical Baselines Are Tuned as Carefully as the Quantum Model, Does Quantum Reservoir Computing Still Win? ​

Author: Tushar Pandey
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.09905v1 Announce Type: cross Abstract: Can a small quantum computer forecast a changing signal better than an ordinary classical method? Many studies say yes, but the classical methods they compare against are often left in a basic, untuned state while the quantum model is carefully optim...

📖 Read original article


199. Depth-Efficient Quantum Topological Data Analysis for Regime-Specific Detection of Financial Stress ​

Author: Arul Rhik Mazumder, Shreyan Ronit Mazumder
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, q-fin.ST

arXiv:2607.09906v1 Announce Type: cross Abstract: We present, to our knowledge, the first adaptation of Pauli Correlation Encoding (PCE) to quantum topological data analysis, reformulating Betti number estimation as a depth-efficient variational optimization over a compressed qubit register. From a ...

📖 Read original article


Author: Sanjeev Khanna, Ashwin Padaki, Erik Waingarten
Published: 7/14/2026, 4:00:00 AM
Categories: cs.DS, cs.LG

arXiv:2607.09909v1 Announce Type: cross Abstract: We study nearest neighbor search from the perspective of data-driven algorithm design: given a dataset $P \subset \mathbb{R}^d$ of size $n$ and sample access to a query distribution over $\mathbb{R}^d$, the goal is to learn a data structure optimized...

📖 Read original article


201. Artificial Intelligence Across the Cardiac Amyloidosis Diagnostic Pathway: From Single-Modality Detection to Multimodal Clinical Integration ​

Author: Diana Shadibaeva, Rochak Dhakal, Kui Zhang, Xiaofeng Yang, Saurabh Malhotra, Weihua Zhou
Published: 7/14/2026, 4:00:00 AM
Categories: physics.med-ph, cs.LG

arXiv:2607.09948v1 Announce Type: cross Abstract: Cardiac amyloidosis (CA) is increasingly recognized but remains substantially underdiagnosed, because its clinical and imaging phenotype overlaps with more common cardiomyopathies. Definitive subtype assignment and management further require integrat...

📖 Read original article


202. A Knowledge-Based Multi-Agent Framework for Security Control Recommendation ​

Author: Carolina Fern'andez-Mart'inez, Shuaib Siddiqui, Vanesa Daza
Published: 7/14/2026, 4:00:00 AM
Categories: cs.GT, cs.AI, cs.CR, cs.LG, cs.MA

arXiv:2607.09954v1 Announce Type: cross Abstract: Hardening IT on-premises environments can be a daunting task for teams without access to adequate cybersecurity expertise. In this regard, Decision Support Systems (DSS) with embedded expert knowledge can assist users by guiding them with security re...

📖 Read original article


203. Inverse-IMPRESSION: A Graph-based Platform for Molecular Structure Elucidation from Experimental NMR Spectroscopic Properties ​

Author: Zheqi Jin, Grace Armitage, Richard Cox, Ben Honor'e, Mohammad Golbabaee, Craig Butts
Published: 7/14/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG

arXiv:2607.09978v1 Announce Type: cross Abstract: Here, we present a platform built on our inverted Graph Transformer Network, IMPRESSION-G2, which can accurately and rapidly reconstruct molecular bonding directly from experimental nuclear magnetic resonance (NMR) spectroscopic information. It compr...

📖 Read original article


204. Beyond Bayesian Nash: Learning Minimax-Regret Equilibria for Adversarial Team Games under Asymmetric Information ​

Author: Naman Aggarwal, Jonathan P. How
Published: 7/14/2026, 4:00:00 AM
Categories: cs.GT, cs.AI, cs.LG, cs.MA

arXiv:2607.09993v1 Announce Type: cross Abstract: Adversarial team games (ATGs) with asymmetric information, such as adversarial path-finding, goal search, and reachability games on graphs, require strategies that are robust to hidden opponent types, such as a hidden goal flag, and to deception. Und...

📖 Read original article


205. Silent Failures in Quantized LLM Reasoning: A Taxonomy-Based Analysis of Hollow Convergence and Failure Mode Shifts ​

Author: Renuka Oladri, Mohan Vamsi Varadaraju Priya, Jerry Wu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.09999v1 Announce Type: cross Abstract: We show that post-training quantization can silently alter how large language models reason even when task accuracy is preserved. Using a six-category failure taxonomy validated by two independent human annotators (Cohen's $\kappa$ = 0.906), we class...

📖 Read original article


206. Manifold Constrained Conformal Prediction for Spatial Events ​

Author: Collin Nill, Trevor Harris, Jason Adams
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2607.10008v1 Announce Type: cross Abstract: We introduce a new conformal prediction method that constructs calibrated prediction sets over collections of spatial events, such as tropical cyclone genesis and earthquake locations. Forecasting natural hazards has become increasingly important, du...

📖 Read original article


207. Runtime Safety Filtering for Learned Small UAS Separation Policies under GNSS Degradation ​

Author: Alex Zongo, Peng Wei
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.MA, cs.SY, eess.SY

arXiv:2607.10014v1 Announce Type: cross Abstract: Learning-based separation assurance for small Unmanned Aircraft Systems (sUAS) achieves near-zero collision rates in simulation, but assumes accurate position and velocity information from Global Navigation Satellite Systems (GNSS). This assumption f...

📖 Read original article


208. Tokenizing Numerical and Embedding Features for LLM RecSys ​

Author: Zhe Xu, Ankit Peshin, Chiyu Zhang, Feng Qi, Johnson Lui, Anil Ramakrishna, Justin Johnson, Carl Hu, Kaushik Rangadurai, Luke Simon
Published: 7/14/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.10016v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as backbone architectures for recommender systems because of their strong sequence modeling and representation learning capabilities. However, most LLM-based recommenders operate primarily on discret...

📖 Read original article


209. A Symbolic Neural CPU for Quantization-Simulated Writeback and Interpretable Program Execution ​

Author: Jose Luis Lima de Jesus Silva
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.AR, cs.LG, cs.NE

arXiv:2607.10021v1 Announce Type: cross Abstract: Neural networks can learn algorithmic input-output mappings, but trusting a learned executor requires more than a correct final answer because the state transitions that produce it are usually hidden. To make those transitions visible, we introduce a...

📖 Read original article


210. Local Multimodal Music Alignment from Global Supervision ​

Author: Irmak Bukey, Zachary Novack, Jongmin Jung, Dasaem Jeong, Chris Donahue
Published: 7/14/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, cs.MM

arXiv:2607.10023v1 Announce Type: cross Abstract: Understanding music requires understanding localized relationships across data modalities, e.g., how time in performance audio maps onto position in a score image. Yet supervision for such local correspondences is difficult to obtain-in practice, we ...

📖 Read original article


211. Robustly Invertible Nonlinear Dynamics and the BiLipREN: From Inversion-Based Control to Generative Trajectory Modelling ​

Author: Yurui Zhang, Ruigang Wang, Ian R. Manchester
Published: 7/14/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC

arXiv:2607.10026v1 Announce Type: cross Abstract: This paper proposes a new notion of robust invertibility for nonlinear dynamical systems, and introduces constructive parameterizations of recurrent neural network which are robustly invertible by design. We define robust invertibility as the existen...

📖 Read original article


212. Are We Ready for AI-Driven Discovery? AI Verification Before the Next Fundamental Physics Breakthrough ​

Author: Gaia Grosso, Vinicius Mikuni, Lukas Heinrich
Published: 7/14/2026, 4:00:00 AM
Categories: physics.data-an, cs.LG, hep-ph

arXiv:2607.10039v1 Announce Type: cross Abstract: Machine learning (ML) has become integral to fundamental physics, accelerating statistical workflows from data acquisition through inference and hypothesis testing. As ML systems grow increasingly autonomous, ensuring their reliability for discovery ...

📖 Read original article


213. DynaFilter: Cloud-driven Dynamic Filtering for Satellite Edge Intelligence ​

Author: Ziyang Zhang, Jie Liu, Luca Mottola
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.10098v1 Announce Type: cross Abstract: Modern satellite edge systems, including those performing remote sensing tasks such object detection and tracking, are characterized by severely limited bandwidth and intermittent connections, making continuous data transmission to the cloud impracti...

📖 Read original article


214. Adaptive Model Compression (AMC): Saliency-Driven Resource Allocation for Ultra-Low-Power Transformer Inference ​

Author: Jiayin Hu, Kai Yuan, Vanessa Hu, Xuetao Yin, Jianhua Li, Sean Suchter
Published: 7/14/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.AR, cs.LG

arXiv:2607.10109v1 Announce Type: cross Abstract: Deploying large-scale transformer models on resource-constrained edge devices remains a challenge due to the high energy and memory overhead inherent in static inference, which processes simple and complex tokens with uniform intensity. To address th...

📖 Read original article


215. Cost of Reasoning in non-English Languages: A Case Study on Japanese ​

Author: Yuu Jinnai
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.10114v1 Announce Type: cross Abstract: Reasoning Language Models (RLMs) achieve their strongest performance when they reason in English, the language for which reasoning-oriented training data is most abundant. However, reasoning trace is a clue for model interpretability and safety, and ...

📖 Read original article


216. On the Efficiency of LoRA Fine-Tuning for Vision-Language-Action Models in Industrial Robotic Manipulation ​

Author: Finn Ferchau, Daniel Pommer, Cristian Axenie
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.10172v1 Announce Type: cross Abstract: Deploying billion-parameter Vision-Language-Action (VLA) models on industrial hardware requires fine-tuning to bridge the embodiment gap. Full Fine-Tuning (FFT) provides maximal plasticity but requires data centre-grade GPUs. We present a systematic ...

📖 Read original article


217. Automated Tensor Scheduling for Hybrid CPU-GPU LLM Inference on Consumer Devices ​

Author: Yangyijian Liu, Hongyi Ye, Mingyang Li, Wu-jun Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG

arXiv:2607.10183v2 Announce Type: cross Abstract: Running large language models on consumer devices such as laptops and desktops is challenging because model weights often exceed GPU memory capacity, making offloading inference necessary to extend effective model capacity with CPU memory. Existing o...

📖 Read original article


218. BiLoG-Net: A Bi-Context Location-Guided Network for Breast Mass Segmentation and Malignancy Classification in Mammography ​

Author: Abu Fatema Mohammad Abdun Noor, Md Imam Ahasan, Md Samiul Ahasan, Kah Ong Michael Goh, S M Hasan Mahmud, Raihana Zannat
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.10188v1 Announce Type: cross Abstract: Breast cancer remains the most commonly diagnosed malignancy among women worldwide, yet accurate detection and characterization of breast masses in mammography remain challenging due to subtle intensity variations, heterogeneous tissue densities, and...

📖 Read original article


219. How much Data do We Need? Sequential Data Collection for Stochastic Programming ​

Author: Xin Li, Juergen Branke, Xuan Vinh Doan
Published: 7/14/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2607.10207v1 Announce Type: cross Abstract: Data-driven optimization often requires collecting data to estimate uncertain model parameters before solving the underlying decision problem. In practice, however, data acquisition may incur non-negligible costs, making it critical to determine when...

📖 Read original article


220. MeloBottleneck: Self-Supervised Melody Skeleton Extraction with a Latent Subsequence Bottleneck ​

Author: Fan Bu, Rongfeng Li, Linfeng Fan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.SD, cs.LG

arXiv:2607.10233v1 Announce Type: cross Abstract: Melody skeleton extraction aims to derive a shorter melody that preserves structural notes while removing ornaments. Prior methods rely on hand-crafted reduction rules or note-wise salience classifiers trained with heuristically or procedurally gener...

📖 Read original article


221. One mechanism for many mental spaces: a shared router over a value slot in language models ​

Author: Oliver Steele, Jiangtao Wen, Yuxing Han
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.10248v1 Announce Type: cross Abstract: Language builds discourse contexts other than the actual: a painting, a belief, a memory, a hypothetical. Each is a mental space in which the same entity can take a different value, as when a flower is red in reality but purple in a portrait. Formal ...

📖 Read original article


222. One Token Is Enough: Fingerprinting and Verifying Large Language Models from Single-Token Output Distributions ​

Author: Tomas Bruckner
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG

arXiv:2607.10252v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly consumed through opaque serving chains - API aggregators, resellers, and inference providers - in which the client has no technical means to confirm that the model answering is the model advertised, and r...

📖 Read original article


223. Byzantine Accountability Without Consensus: Strong Eventual Consistency for Non-Associative, Stochastic, Robust Aggregation ​

Author: Ryan Gillespie
Published: 7/14/2026, 4:00:00 AM
Categories: cs.DC, cs.CR, cs.LG

arXiv:2607.10305v1 Announce Type: cross Abstract: Byzantine-robust aggregation rules such as multi-Krum assume a central coordinator, and decentralising them is obstructed by the rules themselves: they are globally coupled, non-associative, and discontinuous, so an ulpscale perturbation can flip the...

📖 Read original article


224. Gradient-Skipping Relevance Propagation for Efficient Explainability of Vision Transformers ​

Author: Christopher Buratti, Michele Marchetti, Federica Parlapiano, Davide Traini, Domenico Ursino, Luca Virgili
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.10365v1 Announce Type: cross Abstract: Vision Transformers (ViTs) are difficult to interpret because current methods of relevance propagation and attention flow do not fully consider some key architectural features, such as the uneven importance of attention heads and residual connections...

📖 Read original article


225. Stateful Worlds, Stateless Elasticity: Exact-State Serving for Interactive World Models ​

Author: Jin Li (Harvard University), Jiawei Chen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2607.10389v1 Announce Type: cross Abstract: A persistent interactive world model keeps its running state resident on the GPU that serves it: a multi-gigabyte attention cache, almost all of it rewritten at every generation step. That state cannot be recomputed in interactive time or approximate...

📖 Read original article


226. Vertical Fusion: Condensing Internal Representations for Robust ViT Classification ​

Author: Francesco Di Salvo, Shyam Nandan Rai, Hamed Damirchi, Ignacio Meza De la Jara, Sebastian Doerrich, Marco Lents, Christian Ledig
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV

arXiv:2607.10391v1 Announce Type: cross Abstract: Despite exposing rich intermediate representations, Vision Transformers (ViTs) are almost exclusively utilized as black-box feature extractors, where only the last layer is considered for downstream tasks. We challenge this convention by introducing ...

📖 Read original article


227. TVT-PAPD: Pathology-Aware Prototype Distillation for Self-Supervised Whole Slide Image Classification ​

Author: Ramesh Naidu Laveti, Jaya Sreevalsan-Nair, T K Srikanth
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.10406v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has emerged as an effective paradigm for learning transferable representations from large-scale unlabeled whole slide images (WSIs). However, existing SSL methods primarily learn generic visual features and often fail t...

📖 Read original article


228. TSCoNet: A Two-Stage Copula CNN-LSTM for Uncertainty-Aware Spatio-Temporal Forecasting ​

Author: Jongwook Kim, Jong-Min Kim
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.10410v1 Announce Type: cross Abstract: Reliable forecasting of several interrelated environmental variables - such as regional precipitation and temperature, or other correlated geophysical fields - across many locations calls for accurate predictions accompanied by trustworthy statements...

📖 Read original article


229. SPORT: Structure-Aware Prototype Disentanglement for Incomplete Multi-View Clustering ​

Author: Yaoyuan Guo, Zhibin Gu, Songhe Feng, Yuhui Zheng, Bing Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.10413v1 Announce Type: cross Abstract: Prototype-based Incomplete Multi-view Clustering has recently attracted increasing attention by exploiting prototypes as semantic anchors for missing-view imputation. However, existing approaches are still limited in three aspects. First, they typica...

📖 Read original article


230. Is Model Instability just Noise to be Tolerated or a Property that can be Managed? ​

Author: Amirali Rayegan, Lunxiao Li, Tim Menzies
Published: 7/14/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2607.10420v1 Announce Type: cross Abstract: In software analytics, rerunning the same analysis twice often yields different models and conclusions. This reduces trust in the model and limits its use. We find that model instability is a major problem. Across 127 multi-objective SE optimization ...

📖 Read original article


231. Emergent Generalization by Representation Learning in Artificial Neural Networks ​

Author: Hardik Rajpal, Dan Goodman
Published: 7/14/2026, 4:00:00 AM
Categories: q-bio.NC, cs.IT, cs.LG, cs.NE, math.IT, nlin.CD

arXiv:2607.10430v1 Announce Type: cross Abstract: Dimensionality reduction has proven powerful for identifying neural manifolds, which are low-dimensional structures underlying high-dimensional neural activity. These low-dimensional representations have improved the interpretability of population-le...

📖 Read original article


232. Integrating Background Knowledge for Scalable Causal Discovery ​

Author: M'aty'as Schubert, Theofanis Aslanidis, Tom Claassen, Sara Magliacane
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.10456v1 Announce Type: cross Abstract: Expert background knowledge is often available in practical applications of causal discovery. Such constraints on the true causal graph can help causal discovery in terms of identifiability of causal effects and accuracy of the learned structure, but...

📖 Read original article


233. Toward Production-Ready Federated Learning in Healthcare: Privacy, Orchestration, and Governance in MLOps ​

Author: Sakshi Gorkhali, Jonesh Shrestha
Published: 7/14/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2607.10467v2 Announce Type: cross Abstract: Healthcare organizations often cannot freely centralize patient data because medical records are sensitive, regulated, and institutionally controlled. Federated learning offers a practical alternative by allowing hospitals and clinics to train a shar...

📖 Read original article


234. Hallucination Detection in Large Language Models Using Diversion Decoding ​

Author: Basel Abdeen, S M Tahmid Siddiqui, Meah Tahmeed Ahmed, Anoop Singhal, Latifur Khan, Punya Parag Modi, Ehab Al-Shaer
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.10476v1 Announce Type: cross Abstract: Large language models (LLMs) have emerged as a powerful tool for retrieving knowledge through seamless, human-like interactions. Despite their advanced text generation capabilities, LLMs exhibit hallucination tendencies, where they generate factually...

📖 Read original article


235. Fast Data-Driven Modeling of Hydraulic Clutch Control Pressure with Latch-State Classification and Gaussian Process Regression ​

Author: Yash Bagla, Jason Schneider
Published: 7/14/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC

arXiv:2607.10477v1 Announce Type: cross Abstract: This paper presents a data-driven method for modeling the pressure response of a hydraulic clutch control circuit. The system consists of a variable-force solenoid, accumulator, pressure regulator valve, and latch valve, and exhibits nonlinear behavi...

📖 Read original article


236. NetInjectBench: Benchmarking Indirect Prompt Injection in Tool-Using Large Language Model Agents for Network Operations ​

Author: Ruksat Khan Shayoni, Muhammad Faraz Shoaib, S M Asif Hossain, M. F. Mridha
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2607.10490v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents are attractive for network operations, but tickets, alerts, logs, runbooks, and ChatOps messages can carry indirect prompt injections. We present NetInjectBench, a 130-scenario benchmark that separates unt...

📖 Read original article


237. Cross-Layer Misalignment Detection in Agent Skills: A Progressive Loading-Aware Contrastive Learning Approach ​

Author: Chengjun Zhang, Yang Gao, Jianna Hur, Jingjing Zhang, Sagar Samtani
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG

arXiv:2607.10534v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly extended through Agent Skills, reusable artifacts that package natural-language metadata, procedural instructions, and execution-time resources for runtime use. As open-source skill marketplaces expa...

📖 Read original article


238. Representation Learning for Semiparametric Causal Mediation Analysis under No Essential Heterogeneity ​

Author: Roberto Faleh, Sofia Morelli, Holger Brandt
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.10540v1 Announce Type: cross Abstract: We propose a two-stage estimator for structural mediation parameters that combines deep representation learning with G-estimation under the "no essential heterogeneity" (NEH) assumption. We call the method UNIT. In the first stage,TARNet estimates th...

📖 Read original article


239. RecRec: Recursive Refinement for Sequential Recommendation ​

Author: Pervez Shaik, Prosenjit Biswas, Abhinav Thorat, Ravi Kolla, Niranjan Pedanekar
Published: 7/14/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.10541v2 Announce Type: cross Abstract: Sequential recommender systems typically infer user preferences through single-pass encoding of interaction histories without iterative refinement, relying on increasingly deep architectures to capture complex patterns. In this work, we revisit seque...

📖 Read original article


240. Beyond Looking Up, Try Looking Around: Harmonizing Global Structure and Local Consistency in Optimal Transport for Short Text Clustering ​

Author: Zhihao Yao, Yuxuan Gu, Jixuan Yin, Bo Li
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.10548v1 Announce Type: cross Abstract: Pseudo-labeling based on Optimal Transport (OT) has become an effective mechanism for enhancing short text clustering. Existing OT methods are short in modeling semantic consistencies between samples, which may assign different pseudo-labels to seman...

📖 Read original article


241. Projection-Domain Sensitivity Analysis of Vertebral DRRs Under Intrinsic Calibration Perturbation ​

Author: Lin Li, Chaochao Zhou, Benjamin Aubert, Junlin Guo, Junchao Zhu
Published: 7/14/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, cs.NA, math.NA, physics.comp-ph

arXiv:2607.10551v1 Announce Type: cross Abstract: Accurate geometric calibration is essential for fluoroscopy-guided spinal imaging, digitally reconstructed radiograph (DRR) generation, and 2D--3D vertebral registration. Although calibration quality is typically evaluated using reconstruction-based ...

📖 Read original article


242. Observation-Level Watermarking and Detection for Tabular Data ​

Author: Dongyu Cui, Xuan Bi
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ME, cs.LG

arXiv:2607.10554v1 Announce Type: cross Abstract: With the development of generative AI, watermarking techniques have been widely used to detect the authenticity of AI-generated data and protect the rights of users and creators. While it is already well applied in data types including imaging and te...

📖 Read original article


243. BucketKD: A Safety-Aware Bucket-Based Knowledge Distillation Framework for End-to-End Motion Planning ​

Author: Md Nahidul Islam, Mohd Hasan Ali, Dipankar Dasgupta, Myounggyu Won
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.10565v1 Announce Type: cross Abstract: End-to-end motion planning has emerged as a promising paradigm in autonomous driving, directly mapping raw sensor data to control commands via deep neural networks. Despite its advantages, its large model size hinders deployment in resource-constrain...

📖 Read original article


244. When Does Restricting a Coding Agent to execute_code Help? A Regime $\times$ Agent-Design Ablation ​

Author: Hong Yang, Qi Yu, Travis Desell
Published: 7/14/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2607.10569v1 Announce Type: cross Abstract: Modern coding agents expose multiple tool surfaces -- IDE primitives, bash, and Model Context Protocol (MCP) code-execution -- and the field has shipped three contradictory claims about which one matters. We run the missing crossed comparison: an int...

📖 Read original article


245. Approximation of Analytic Functions by ReLU Neural Networks with Adjustable Depth and Width ​

Author: Yanming Lai, Defeng Sun, Yang Wang
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, cs.NA, math.IT, math.NA

arXiv:2607.10589v1 Announce Type: cross Abstract: In contrast to most studies on neural network approximation theory that characterize results through a single parameter, such as the total number of network parameters, \cite{shen2020deep} pioneered the characterization of approximation rates as a jo...

📖 Read original article


246. End-to-End Real-Time Drone-Based Person Detection Framework Using Deep Learning ​

Author: Payel Sarmah, Ayush Ranjan, Piyush Kaushik Bhattacharyya, Anil Kr. Shaw, Pradip Kr. Das
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.10605v1 Announce Type: cross Abstract: In recent years, Unmanned Aerial Vehicles (UAVs) or drones have gained rapid response in terms of security, search and rescue (SAR), border surveillance, etc. Existing monitoring frameworks often struggle to maintain detection consistency when target...

📖 Read original article


247. Demixing Sparse Signals from Nonlinear Observations using Generalized Non-convex Regularization ​

Author: Raziyeh Takbiri
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, eess.SP

arXiv:2607.10618v1 Announce Type: cross Abstract: We consider the recovery of a pair of sparse vectors from a limited number of nonlinear observations of their superposition: $y_i=g(\inner{\ba_i}{\bPhi\bw^\ast+\bPsi\bz^\ast})+e_i$, $i=1,\dots,m$, with $m\ll n$, incoherent orthonormal bases $\bPhi,\b...

📖 Read original article


248. Learning Topological Quantum Phases from Limited Subsystems ​

Author: Mehran Khosrojerdi, Sougato Bose, Alessandro Cuccoli, Paola Verrucchi, Abolfazl Bayat, Leonardo Banchi
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cond-mat.stat-mech, cond-mat.str-el, cs.LG

arXiv:2607.10656v1 Announce Type: cross Abstract: Characterizing quantum topological phases requires measuring non-local string order parameters, demanding access to the full system, which is often experimentally unfeasible. In this work, we introduce a data-efficient supervised learning framework t...

📖 Read original article


249. Edge Cluster Expansion with Radial Rotary Attention for Interatomic Potentials ​

Author: Zemin Xu, Wenbo Xie, P. Hu
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.mtrl-sci, cs.LG, physics.chem-ph

arXiv:2607.10664v1 Announce Type: cross Abstract: In this paper, we provide a systematic investigation of SO(2) theory to machine learning interatomic potentials (MLIPs) and identify the limitations of conventional SO(2) Linear architectures relative to SO(3) Clebsch-Gordan Tensor Products (CGTP). B...

📖 Read original article


250. Answer-Conditioned Chain-of-Thought Distillation for Few-Shot Industrial Vision with Small VLMs ​

Author: Shubham Rao
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cond-mat.mtrl-sci, cs.AI, cs.LG

arXiv:2607.10666v1 Announce Type: cross Abstract: Deploying AI-based visual inspection in manufacturing is hard because requirements change often, new defect types appear, and large labeled datasets are rarely available. We propose answer-conditioned chain-of-thought (CoT) distillation for rapidly a...

📖 Read original article


251. Action Map Policy: Learning 3D Closed-loop Manipulation via Pixel Classification ​

Author: Haojie Huang, Zhang Ye, Linfeng Zhao, Boce Hu, Mingxi Jia, Yu Qi, Ahmed Agha, Dian Wang, Robert Platt, Robin Walters
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG

arXiv:2607.10706v1 Announce Type: cross Abstract: The action space poses a major challenge in robot learning, since it is often high-dimensional, can span long time horizons, and frequently admits multi-modal optimal solutions. A good choice of action representation and loss function can help to add...

📖 Read original article


252. MDQEC-QAS: Meta-Decoding for Quantum Error Correction with Hardware-Aware VQC Search and Confidence-Gated Recovery ​

Author: Prashant Kumar Choudhary, Nouhaila Innan, Muhammad Shafique, Rajeev Singh
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.AR, cs.LG

arXiv:2607.10707v1 Announce Type: cross Abstract: We propose a unified meta-decoding framework for quantum error correction that learns syndrome-to-recovery mappings across multiple stabilizer codes and noise settings, without requiring separate decoders for each configuration. The benchmark include...

📖 Read original article


253. WattCouncil: Context-Aware Household Energy Scenario Generation With Governed LLMs ​

Author: Mohannad Takrouri, Nicolas M. Cuadrado A., Martin Tak'a\v{c}
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2607.10720v1 Announce Type: cross Abstract: The accelerating shift toward low-carbon power systems, together with the widespread adoption of behind-the-meter technologies such as rooftop solar and electric vehicles, is placing new operational and analytical demands on electricity grids. At the...

📖 Read original article


254. Filtering Harmful Actions Isn't Enough: Phantom Transfer in Agentic SDF ​

Author: Chinmayi Dixit
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.10750v1 Announce Type: cross Abstract: Synthetic data is widely used to train large language models because it is inexpensive to generate and easy to control. As models are increasingly deployed as agents, synthetic trajectories are likely to become an important source of training data fo...

📖 Read original article


255. TOLiD: Bridging the Architecture Gap in Vision Foundation Model to LiDAR Pretraining via Token Lifting for Distillation ​

Author: Sutharsan Mahendran, Darshana Priyasad, Kaushik Roy, Tharindu Fernando, Sridha Sridharan, Clinton Fookes, Peyman Moghadam
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO

arXiv:2607.10762v1 Announce Type: cross Abstract: Cross-modal distillation from Vision Foundation Models (VFMs) to LiDAR backbones has recently emerged as a self-supervised pretraining strategy that reduces reliance on dense point-wise annotation for 3D scene understanding. However, existing distill...

📖 Read original article


256. Lightning Fast Matching Dependency Discovery with Desbordante ​

Author: Alexey Shlyonskikh, Michael Sinelnikov, Daniil Nikolaev, Yurii Litvinov, George Chernishev
Published: 7/14/2026, 4:00:00 AM
Categories: cs.DB, cs.AI, cs.LG, cs.PF

arXiv:2607.10771v1 Announce Type: cross Abstract: Matching dependency is a generalization of the functional dependency concept, which allows users to apply custom similarity functions for matching individual attributes. Matching dependencies have a wide range of applications for solving various data...

📖 Read original article


257. Toward Efficient Weakly Supervised Semantic Segmentation Using Only Low-Magnification Histopathological Images ​

Author: Dung Minh Do, Nhat-Thanh Huynh, Duc Minh Huynh, Doanh C. Bui, Khang Nguyen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.10783v1 Announce Type: cross Abstract: Whole-slide images (WSIs) provide rich tissue-level and cellular-level information, but storing and transmitting high-magnification pathology data is resource-intensive. Moreover, annotating WSIs at the pixel level is labor-intensive and time-consumi...

📖 Read original article


258. Q-Learning Lab: Teaching Reinforcement Learning Through Learner-Generated Trace Analysis ​

Author: Ekkachai Jueng
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CY, cs.LG

arXiv:2607.10802v1 Announce Type: cross Abstract: Reinforcement learning is usually introduced through the Bellman update, yet the equation often remains abstract to undergraduates: they watch policy arrows converge but rarely observe how each value is computed or why an action is chosen. We present...

📖 Read original article


259. Diagnosing and Mitigating Thinking Collapse in On-Policy Self-Distillation ​

Author: Keqin Peng, Chen Li, Yuanxin Ouyang, Yancheng Yuan, Liang Ding
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.10805v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) has emerged as a crucial paradigm for enhancing and aligning Large Language Models (LLMs). However, in complex reasoning tasks, OPSD paradoxically degrades downstream performance. In this paper, we systematically in...

📖 Read original article


260. Large Language Models for Token-Efficient and Semantic-Preserving Opinion Summarization ​

Author: Fabrizio Marozzo, Stefano Iannicelli
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.10825v1 Announce Type: cross Abstract: Opinionated text - spanning product reviews, hotel feedback, and social posts - captures rich signals about user experiences, preferences, and concerns. However, the scale, redundancy, and imbalance of such corpora make it challenging to analyze opin...

📖 Read original article


261. Route, Communicate, and Reason: Gated Routing and Adaptive Depth for Efficient Multi-Agent Reasoning ​

Author: Sudipto Ghosh, Tanmoy Chakraborty
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.10836v1 Announce Type: cross Abstract: Multi-agent ensembling multiplies active parameters and inference cost without answering three basic questions: which agents to consult, how deeply a query should traverse a hierarchy of agents, and when inter-agent communication is worth its cost. W...

📖 Read original article


262. Diversify Diffusion with Temperature Sampling and Variance-Corrective Time Shifting ​

Author: Peizhuo Li, Emre Aksan, Alexandru-Eugen Ichim, Thabo Beeler, Olga Sorkine-Hornung
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.10853v1 Announce Type: cross Abstract: Diffusion models faithfully reproduce their training distribution, but also inherit its imbalances and leave rare or under-represented modes hard to reach. A natural inference-time remedy is to sample from the high-temperature target $p^{(\gamma)}_0(...

📖 Read original article


263. Transferable Implicit Solvent Machine Learning Potential for Drugs and Proteins Approaching Ab Initio Accuracy ​

Author: Jan Eckwert, Julija Zavadlav
Published: 7/14/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, q-bio.BM

arXiv:2607.10887v1 Announce Type: cross Abstract: Machine learning interatomic potentials (MLPs) have revolutionized atomistic modeling, offering the potential to replace traditional methods like Density Functional Theory (DFT). However, inference time of MLPs is orders of magnitude slower than that...

📖 Read original article


264. ZoRRO: A Zero-Weight Personalized Recommender System for Scalable News Recommendation ​

Author: Johannes Kruse, Ryotaro Shimizu, Kasper Lindskow, Jon Tofteskov, Michael Riis Andersen, Julian McAuley, Jes Frellsen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.10910v1 Announce Type: cross Abstract: We present ZoRRO (Zero-Weight Personalized Recommender System), a zero-weight, training-free framework for personalized news recommendation designed for scalable real-world deployment. ZoRRO outperforms strong neural baselines in offline ranking eval...

📖 Read original article


265. Normative Alignment of Recommender Systems via Internal Label Shift ​

Author: Johannes Kruse, Kasper Lindskow, Michael Riis Andersen, Ryotaro Shimizu, Julian McAuley, Pierre-Alexandre Mattei, Jes Frellsen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.10915v1 Announce Type: cross Abstract: We introduce NAILS (Normative Alignment of Recommender Systems via Internal Label Shift), a simple and scalable method for aligning recommendation outputs with target distributions over item-level attributes, such as categories. Recommender systems o...

📖 Read original article


266. Fast Whole-Brain, Geometry-Aware Functional Alignment for Cross-Subject Decoding ​

Author: Pierre-Louis Barbarant, Florent Meyniel, Bertrand Thirion
Published: 7/14/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG, stat.ML

arXiv:2607.10931v1 Announce Type: cross Abstract: Decoding brain activity is useful for characterizing brain processes and understanding the functional architecture underlying cognition. However, the inter-individual variability in brain response patterns limits the development of decoders that gene...

📖 Read original article


267. CGS: Configurable Graph Summarization with Bounded Neighborhood Loss and Query Support ​

Author: Shubhadip Mitra, Sona Elza Simon, C Oswald, Arnab Bhattacharya, Arindam Pal
Published: 7/14/2026, 4:00:00 AM
Categories: cs.DS, cs.AI, cs.LG

arXiv:2607.10969v1 Announce Type: cross Abstract: Given a large graph, how to generate a compact summary graph that is configurable by the user and supports multiple graph queries with either no loss or with high accuracy? The ever growing size of graph datasets makes the above question on graph sum...

📖 Read original article


268. EquiFusion: Kinematics-Agnostic Human Motion Prediction via Equivariant Latent Diffusion ​

Author: Cecilia Curreli, Florian Hofherr, Dominik Muhle, Abhishek Saroha, Riccardo Marin, Daniel Cremers
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.HC, cs.LG

arXiv:2607.10984v1 Announce Type: cross Abstract: Existing Stochastic 3D Human Motion Prediction models are fundamentally constrained by hard-coding the skeleton kinematics, severely limiting generalization, preventing cross-dataset training, and requiring complex data retargeting. We introduce Equi...

📖 Read original article


269. Actor-Critic Learning for Extended Mean Field Control with Deterministic Policies ​

Author: Ziheng Cheng, Xin Guo, Huy^en Pham, Yufei Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: math.OC, cs.AI, cs.LG, stat.ML

arXiv:2607.11005v1 Announce Type: cross Abstract: This paper develops a model-free reinforcement learning framework for continuous--time extended mean field control problems, where both the dynamics and reward may depend on the joint distribution of states and controls. We adopt deterministic feedba...

📖 Read original article


270. Overcoming Fourier Locking in Quantum Data Re-uploading Classifiers via Spectral Homotopy ​

Author: Spencer Topel
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.11013v1 Announce Type: cross Abstract: Data re-uploading parameterized quantum circuits (DRU-PQCs) are universal function approximators, yet their expressivity produces oscillatory, non-convex loss landscapes that resist gradient-based optimization. We show that the primary optimization b...

📖 Read original article


271. Can a Language Model Learn Facts Continually in Its Weights? ​

Author: Charles O'Neill
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.11020v2 Announce Type: cross Abstract: Continual learning promises a language model that keeps acquiring knowledge after training, with each new fact written into its weights. Whether weight writes can support accumulation remains undecided. We follow invented facts written into Qwen3 mod...

📖 Read original article


272. Reference-Based Face Super-Resolution Using the Spatial Transformer ​

Author: Varun Ramesh Jois, Antonella DiLillo, James Storer
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.11025v1 Announce Type: cross Abstract: Face super-resolution is the task of increasing the resolution of an image containing a face thereby adding finer detail. It is a ubiquitous task in many computer vision applications and quite often the user isn't even aware that it is being performe...

📖 Read original article


Author: Zhen-Lin Chen, Maosen Sheng, Peng Lin, Jianmin Chen, Zhuojian Xiao, Dongyue Wang, Xiwei Zhao
Published: 7/14/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, cs.MM

arXiv:2607.11030v1 Announce Type: cross Abstract: Multimodal information is pivotal for e-commerce search ranking. Existing works leverage multimodal data typically by fine-tuning general Multimodal Large Language Models (MLLMs) via collaborative signals, subsequently integrating the derived represe...

📖 Read original article


274. When cheap gradients fail: the measurement cost of attacking quantum classifiers ​

Author: Bacui Li, Chandra Thapa, Tansu Alpcan, Udaya Parampalli
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.CR, cs.LG

arXiv:2607.11095v1 Announce Type: cross Abstract: Adversarial perturbations threaten machine learning classifiers, including variational quantum classifiers. We show that finite quantum measurement statistics (shot noise) act as a built-in defense against gradient-based test-time attacks whose cost ...

📖 Read original article


275. Implicit Neural Networks as Static Controllers: Certificates and Performance Separation ​

Author: Giuseppe C. Calafiore, Laurent El Ghaoui
Published: 7/14/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2607.11122v1 Announce Type: cross Abstract: Implicit neural controllers (INCs) are static feedback laws that are evaluated through an algebraic fixed point {equation}; they include as special cases neural network controllers. We propose a so-called implicit representation of neural networks as...

📖 Read original article


276. Comparison-Based Ordinal Learning for Proactive Driving Risk Assessment ​

Author: Zhuoren Li, Yi Zhong, Weiqi Zhang, Xinrui Zhang, Lu Xiong, Chongfeng Wei, Bo Leng
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.11128v1 Announce Type: cross Abstract: Real-time driving risk assessment provides an essential basis for proactive safety by identifying and quantifying the danger of ongoing road interactions before adverse outcomes occur. However, due to the scarcity of collision data and frame-level ri...

📖 Read original article


277. A Formal Hierarchical Architecture for Agentic Orchestration with Stack-Based Execution and Lazy Discovery ​

Author: Prashant Devadiga, Abhishek, Adithya Mishra, Alok Singh, Amisha Sinha, Asit Desai, Gaurang Dahad, Harshit Bhushan, Mandati Pramod Reddy, Prakhar Gupta, Rupesh Patil, Siddhi Behere
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.11138v1 Announce Type: cross Abstract: The rapid expansion of capabilities in Large Language Model (LLM) agents has exposed a critical architectural bottleneck: when agents are given access to a flat, monolithic registry of tools, the model must evaluate hundreds or thousands of options s...

📖 Read original article


278. Pix2Act: Image-Space Manipulation Policies with Equivariant Augmentation ​

Author: Haojie Huang, Linfeng Zhao, Haotian Liu, Zhang Ye, Si-Yuan Huang, Mingxi Jia, Boce Hu, Fangzhou Lin, Yu Qi, Dian Wang, Robin Walters, Robert Platt
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.11167v1 Announce Type: cross Abstract: Representing manipulation actions as 2D trajectories in the camera plane provides a compact and interpretable basis for learning complex 3D manipulation policies. However, it also creates challenges from out-of-frame trajectories and limited precisio...

📖 Read original article


279. STAMP: Provenance-Guided Credit Assignment for Deep Search Agents ​

Author: Ke Xu, Han Xu, Xinran Chen, Yuqian Wang, Zhixuan Li, Xiaojian Liu, Changwo Wu, Jianqiang Xia, Yuchen Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.11172v1 Announce Type: cross Abstract: Reinforcement learning for deep-search agents has largely focused on trajectory-level scoring -- outcome correctness, citation-aware rewards, and evidence coverage. Yet the actions that expose supporting documents receive no targeted credit, a gap we...

📖 Read original article


280. PREF-Gate: Provenance-Constrained Relational Evidence Fusion with Validation-Gated Selection for Graph Fraud Detection ​

Author: Liming Liu, Chao Hu, Mingfei Lu, Yiwei Ge, Xingle Li, Heyuan Shi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.11212v1 Announce Type: cross Abstract: Relational fraud detection can exploit both label-free graph context and label-derived neighborhood evidence, but these two information sources obey different validity conditions. In particular, neighborhood risk becomes invalid when a queried node's...

📖 Read original article


281. LaGuadia: Language-Guided Adaptive Distillation from Pathology Foundation Models ​

Author: Gangsu Kim, Won-Ki Jeong
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.11257v1 Announce Type: cross Abstract: Pathology Foundation Models (PFMs) offer powerful Whole Slide Image (WSI) representations but suffer from massive computational costs. While Knowledge Distillation (KD) can create efficient student models, existing multi-teacher methods often use sub...

📖 Read original article


282. Bringing Back Rule Induction to Fluid Intelligence Research? An Initial Validation of the ARC-AGI Benchmark in Humans ​

Author: Jasmin Thelen, Oliver Wilhelm
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.11263v2 Announce Type: cross Abstract: Two competing perspectives on fluid intelligence (gf) measures propose that performance is primarily constrained either by working memory capacity or by the ability to induce novel relations. The first perspective is currently dominant in measurement...

📖 Read original article


283. Long-Memory Reservoir Computing for Data-Scarce Dengue Forecasting ​

Author: Rahul Goswami, Shinjini Paul, Palash Ghosh, Tanujit Chakraborty
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.11272v1 Announce Type: cross Abstract: Accurate dengue forecasting is crucial for public health planning, but remains challenging because incidence series are often short, noisy, non-stationary, nonlinear, and often affected by long-range temporal dependence. Fractional differencing in Au...

📖 Read original article


284. Fixed-Protocol Amortized MPS Tomography with Conformalized Predictive Uncertainty ​

Author: Jian Xu, Delu Zeng, John Paisley, Qibin Zhao
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.11273v1 Announce Type: cross Abstract: Quantum state tomography is sample-starved, and the states one prepares live on a narrow, learnable manifold. A $k{=}0$ prior-only control shows that on concentrated families a prior estimate is already near-optimal, so ``high fidelity at few measure...

📖 Read original article


285. Backpropagation as a Nilpotent Linear System ​

Author: Ahmed Boughammoura
Published: 7/14/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2607.11289v1 Announce Type: cross Abstract: Backpropagation is the computational engine of deep learning, yet its mathematical structure is typically treated as a procedural traversal of computational graphs. We present a global operator theory of the \emph{F-adjoint} framework, which reformul...

📖 Read original article


286. Inter-Stop Energy Prediction and Causal Driver Quantification for Dual-Source Trolleybuses via a Time-Aware Tabular Deep Learning Architecture ​

Author: Wentao Zeng (School of Management, Foshan University, Foshan, China a School of Management, Foshan University, Foshan, China, School of Mechanical and Electrical Engineering and Automation, Foshan University, Foshan, China), Zijian Huang (School of Artificial Intelligence, South China Normal University, Guangzhou, China), Yiming Bie (School of Transportation, Jilin University, Changchun, China), Jiabin Wu (School of Management, Foshan University, Foshan, China a School of Management, Foshan University, Foshan, China), Jun Gong (Department of Civil Engineering, The University of Hong Kong, Hong Kong, China)
Published: 7/14/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2607.11349v1 Announce Type: cross Abstract: Dual-source trolleybuses alternate between overhead catenary supply and on-board battery operation, creating energy-use patterns driven by route attributes, high-frequency trajectories, and hourly weather. Existing models struggle to represent these ...

📖 Read original article


287. Characterising AI Models for Cataloguing ​

Author: Miguel Arana-Catania, Neil Jefferies
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.DL, cs.IR, cs.LG

arXiv:2607.11353v1 Announce Type: cross Abstract: The creation of digital collections involves not only the digitisation of content, but also the creation of catalogue records for it. This often-overlooked task requires slow and costly expert manual work. In this project, we have evaluated the appli...

📖 Read original article


288. Decomposing Runtime, Kernel, and Quantization Speedups via a Matched FP16 Intermediate: A Hardware-Conditioned Case Study on Four NVIDIA RTX A5000 GPUs ​

Author: Weijia Han, Lisha Qu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.DC, cs.LG, cs.PF

arXiv:2607.11368v1 Announce Type: cross Abstract: Reported serving speedups from quantized kernels typically bundle the weight format, the kernel, and the inference runtime into one number. We present an attribution study on four NVIDIA RTX A5000 GPUs, 24 GiB each, on a single host with NVLink-bridg...

📖 Read original article


289. StructAgent: Harness Long-horizon Digital Agents with Unified Causal Structure ​

Author: Wenyi Wu, Sibo Zhu, Kun Zhou, Aayush Salvi, Zixuan Song, Biwei Huang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2607.11388v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) and vision-language models (VLMs) have enabled increasingly capable digital agents for computer use. However, real-world tasks are often long-horizon and involve evolving contexts containing accumulated...

📖 Read original article


290. Climate-Invariant Conformal Prediction Intervals for Multi-Horizon Solar and Wind Forecasting ​

Author: Shreedhar Gangwar (B. R. Ambedkar National Institute of Technology, Jalandhar, India), Abhinav Bains (B. R. Ambedkar National Institute of Technology, Jalandhar, India), Banalaxmi Brahma (B. R. Ambedkar National Institute of Technology, Jalandhar, India)
Published: 7/14/2026, 4:00:00 AM
Categories: stat.AP, cs.LG

arXiv:2607.11470v1 Announce Type: cross Abstract: Reliable uncertainty quantification is essential for integrating solar and wind generation into modern power systems, where operators must weigh risk rather than act on point forecasts alone. Existing probabilistic methods, however, often either lack...

📖 Read original article


291. Towards Efficient Convolutional Neural Network for Embedded Hardware via Multi-Dimensional Pruning ​

Author: Hao Kong, Di Liu, Xiangzhong Luo, Shuo Huai, Ravi Subramaniam, Christian Makaya, Qian Lin, Weichen Liu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.11473v1 Announce Type: cross Abstract: In this paper, we propose TECO, a multi-dimensional pruning framework to collaboratively prune the three dimensions (depth, width, and resolution) of convolutional neural networks (CNNs) for better execution efficiency on embedded hardware. In TECO, ...

📖 Read original article


Author: H. Xu, B. He, S. Wang, Y. Jiang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT

arXiv:2607.11488v1 Announce Type: cross Abstract: Long range-frequency hopping spread spectrum (LR-FHSS) is a promising uplink physical layer for massive low Earth orbit satellite Internet of Things, where low power terminals report short packets from wide area regions with limited terrestrial infra...

📖 Read original article


293. Learning Residual Kinematic Corrections for Continuous Neural Decoding via Reinforcement Learning ​

Author: Jiamian Li, Niall McShane, Attila Korik, Naomi du Bois, Karl McCreadie, Leen Jabban, Benjamin Metcalfe, "Ozg"ur \c{S}im\c{s}ek, Damien Coyle
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.11530v1 Announce Type: cross Abstract: Decoding continuous three-dimensional (3D) motor imagery (MI) using non-invasive electroencephalography (EEG)-based brain--computer interfaces (BCIs) remains challenging due to signal variability and residual decoding errors. Deep learning architectu...

📖 Read original article


294. Adaptive Routing for Efficient Diffusion Transformer-Based PNI Prediction ​

Author: Youngung Han, Dohyun Kweon, Kyeonghun Kim, Hyunsu Go, Jina Jeong, Suah Park, Induk Um, Junga Kim, Anna Jung, Yului Jeong, Sungha Park, Jinyong Jun, Pa Hong, Woo Kyoung Jeong, Won Jae Lee, Ken Ying-Kai Liao, Hyuk-Jae Lee, Nam-Joon Kim
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.11533v1 Announce Type: cross Abstract: Perineural invasion (PNI) is a critical prognostic factor in cholangiocarcinoma. However, its preoperative prediction from magnetic resonance imaging (MRI) remains challenging due to subtle imaging features that extend beyond tumor boundaries into su...

📖 Read original article


295. Tropical Circuits with Scalar Multiplication Gates ​

Author: Christoph Hertrich, Moritz Stargalla
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CC, cs.LG, math.CO

arXiv:2607.11540v1 Announce Type: cross Abstract: We study tropical circuits with scalar multiplication gates, that is, algebraic circuits whose gates implement $\max$, $+$, or multiplication with a positive constant. For such circuits, we prove exponential size lower bounds for computing maximum we...

📖 Read original article


296. Training-Free Off-Screen Player Imputation for Broadcast-Based Spatial Football Analytics ​

Author: Seongjin Choi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.11548v1 Announce Type: cross Abstract: Spatial football metrics such as pitch control assume access to the positions of all 22 players, yet the most widely available source of positional data -- the broadcast main camera -- shows only 10-16 of them at any moment. We quantify the resulting...

📖 Read original article


297. Machine Learning-Based Reconstruction for Resistive Silicon Sensors ​

Author: Alexander Aoki, Gaetano Barone, Leena Diehl, Gabriele Giacomini, Vagelis Gkougkousis, Hanshal Goyal, Rohan Kher, Daniel Li, Anna Macchiolo, Yevhenii Padnuik, Daria Senina, Samantha Sunnarborg, Jessica Tang, Alessandro Tricoli, Lixing Wang, Don C. Wong
Published: 7/14/2026, 4:00:00 AM
Categories: hep-ex, cs.LG, nucl-ex

arXiv:2607.11585v1 Announce Type: cross Abstract: Low-Gain Avalanche Diodes (LGADs) and AC-coupled Low-Gain Avalanche Diodes (AC-LGADs) are promising technologies for precision timing and four-dimensional tracking. In AC-LGADs, the AC pad is coupled to the resistive n$^{+}$ layer through a dielectri...

📖 Read original article


298. Globally Consistent Coloring Schemes for Language Identification ​

Author: Moses Charikar, Jon Kleinberg, Chirag Pabbaraju
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.DS, cs.LG

arXiv:2607.11606v1 Announce Type: cross Abstract: We study how little extra information is needed to make adversarial language learning possible. In Gold's model of language identification in the limit, a learner is given an enumeration of the strings from an unknown language chosen from a countable...

📖 Read original article


299. Auditing the Risk Claims of Distributional Reinforcement Learning ​

Author: Hari Prasad
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, stat.ML

arXiv:2607.11607v1 Announce Type: cross Abstract: Distributional reinforcement learning agents learn full return distributions that are increasingly read at face value: for interpretability, risk-sensitive control, and safety monitoring. We ask a question theory anticipates but that has not been mea...

📖 Read original article


300. Lesioned Multimodal Language Models Reproduce Aphasic Picture-Naming Patterns ​

Author: Yong Yang, Xiang Guan, Sophie Arheix-Parras, Saeed Ahmadi, Roger Newman-Norlund, Leonardo Bonilha, Christopher Rorden, Julius Fridriksson, Rutvik H. Desai, Srihari Nelakuditi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.11621v1 Announce Type: cross Abstract: Aphasia following stroke commonly produces systematic naming errors with characteristic profiles, but whether general-purpose language models not designed for clinical simulation can reproduce these patterns remains untested. We investigated (1) whet...

📖 Read original article


301. SKooP: Symmetric Koopman Predictions for Faster and More Generalizable Legged Robot Locomotion with Reinforcement Learning ​

Author: Evelyn D'Elia, Weishu Zhan, Giulio Turrisi, Giulio Romualdi, Giuseppe L'Erario, Raffaello Camoriano, Wei Pan, Daniele Pucci
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.11624v2 Announce Type: cross Abstract: Reinforcement learning (RL) algorithms classically suffer from poor sample efficiency. In robotics, a recent line of work has emerged addressing this problem by encoding physics priors in the learning process. However, most of these approaches are va...

📖 Read original article


302. Reproducing human biases in route choice using large language models: Toward scalable behavioral modeling ​

Author: Jiangtao Han, Shoufeng Ma, Shuxian Xu, Geng Li, Shuai Ling, Ning Jia, Zhengbing He
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.SI, physics.soc-ph

arXiv:2607.11632v1 Announce Type: cross Abstract: Human choice behavior, including route choice, exhibits systematic behavioral biases that deviate from the assumptions of full rationality. Cumulative prospect theory (CPT) has been widely recognized as an effective framework for characterizing such ...

📖 Read original article


303. Imputation-free transformer learning enables robust Alzheimer's disease prediction and calibrated uncertainty quantification across heterogeneous clinical cohorts ​

Author: Christelle Schneuwly Diaz, Narmina Baghirova, Duy-Thanh Vu, Duy-Cat Can, Gilles Allali, Philippe Ryvlin, Oliver Y. Ch'en
Published: 7/14/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG

arXiv:2607.11656v2 Announce Type: cross Abstract: Accurate diagnostic classification and disease-severity prediction for Alzheimer's disease are hampered by the incompleteness and heterogeneity of real-world clinical data. Left unaddressed, these barriers prevent reliable disease modelling and hinde...

📖 Read original article


304. Diversified Multinomial Logit Contextual Bandits ​

Author: Heesang Ann, Taehyun Hwang, Min-hwan Oh
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.11684v1 Announce Type: cross Abstract: Existing contextual multinomial logit (MNL) bandits model relevance-driven choice but ignore the potential benefits of within-assortment diversity, while submodular/combinatorial bandits encode diversity in rewards but lack structured choice probabil...

📖 Read original article


305. Self-Healing Visual Recovery for Autonomous Ground Vehicles Using Camera-Only Visual Odometry ​

Author: Jakob Solberg Berntzen, Safia Fatima, Leon Moonen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SE

arXiv:2607.11686v1 Announce Type: cross Abstract: Low-cost unmanned ground vehicles are often used in indoor places like warehouses, inspection corridors, and farm rows, where painted floor lines guide the robot. Line following is useful because it only needs one camera and little computing power, b...

📖 Read original article


306. $\mathtt{Q^2SAR}$: overcoming classical bottlenecks in drug discovery via quantum multiple kernel learning ​

Author: Mariano Caruso, Daniel Ruiz, Alejandro Giraldo, Guido Bellomo
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.11701v1 Announce Type: cross Abstract: Quantitative Structure-Activity Relationship ($\mathtt{QSAR}$) modeling is a foundational computational methodology in early-stage drug discovery, heavily relied upon for predicting compound toxicity, bioavailability, and therapeutic potential. Howev...

📖 Read original article


307. NeuralActuator: Neural Actuation Modeling for Robot Dynamics and External Force Perception ​

Author: Zhiyang Dou, John U. Onyemelukwe, Hangxing Zhang, Heng Zhang, Minghao Guo, Yunsheng Tian, Michal Piotr Lipiec, Joshua Jacob, Chao Liu, Peter Yichen Chen, Yuri Ivanov, Wojciech Matusik
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.GR, cs.LG

arXiv:2607.11734v1 Announce Type: cross Abstract: Differentiable simulators have advanced policy learning and model-based control, yet actuator dynamics remain an important source of sim-to-real error. This is particularly acute on low-cost platforms, where the linear current-to-torque relation $\ta...

📖 Read original article


308. When Local Monitors Miss Compositional Harm: Diagnosing Distributed Backdoors in Multi-Agent Systems ​

Author: Yibo Hu, Ren Wang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.MA

arXiv:2607.11751v1 Announce Type: cross Abstract: As multi-agent, tool-using LLM systems are deployed, a common safety net is a runtime monitor that checks each message, tool call, or step on its own. We show this net has a fundamental hole. A distributed backdoor splits a harmful payload across age...

📖 Read original article


309. Paradoxes of Game Theoretic Equilibria and Price of Anarchy ​

Author: Georgios Piliouras, Ian Gemp, Siqi Liu, Luke Marris
Published: 7/14/2026, 4:00:00 AM
Categories: cs.GT, cs.LG, cs.MA, math.DS, math.OC

arXiv:2607.11752v1 Announce Type: cross Abstract: For decades, static solution concepts (Nash, Correlated, and Coarse Correlated Equilibria) and the Price of Anarchy (PoA) have formed the bedrock of algorithmic game theory, with no-regret learning proving fast convergence to such game-theoretic equi...

📖 Read original article


310. Input-Aware Dynamic Backdoor Attack Against Quantum Neural Networks ​

Author: Junrui Zhang, Zemin Chen, Lusi Li, Mohammad Ghasemigol, Daniel Takabi, Rui Ning
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.11843v1 Announce Type: cross Abstract: Quantum Neural Networks (QNNs) are a promising framework for quantum machine learning on near-term quantum devices, but their security risks remain insufficiently understood. Studies have shown that QNNs are vulnerable to backdoor attacks, yet existi...

📖 Read original article


311. A Durability and Cross-Language Transfer Benchmark for a Validated Teaching-Feedback Classification Protocol ​

Author: Esteban U. Vega Barajas
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.11873v1 Announce Type: cross Abstract: Institutions collect far more open-ended teaching-evaluation feedback than they read. A prior study introduced a validated protocol for classifying such comments by thematic category and sentiment, built from a documented annotation guide, an intra-a...

📖 Read original article


312. A Minimalist Retargeting-Guided Reinforcement Learning Recipe for Dexterous Manipulation ​

Author: Yunhai Feng, Natalie Leung, Jiaxuan Wang, Lujie Yang, Haozhi Qi, Preston Culbertson
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.11874v1 Announce Type: cross Abstract: Recent work in humanoid whole-body control has found success with a simple recipe: retarget human motion to robot kinematic references, then train policies via reinforcement learning (RL) to track them. But how does this recipe transfer to dexterous ...

📖 Read original article


313. Learning to Schedule in Parallel-Server Queues with Stochastic Bilinear Rewards ​

Author: Jung-hun Kim, Milan Vojnovic
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, math.OC

arXiv:2112.06362v5 Announce Type: replace Abstract: We consider the problem of scheduling in multi-class, parallel-server queuing systems with uncertain rewards from job-server assignments. In this scenario, jobs incur holding costs while awaiting completion, and job-server assignments yield observa...

📖 Read original article


314. Federated Topic Model and Model Pruning Based on Variational Autoencoder ​

Author: Chengjie Ma, Yawen Li, Meiyu Liang, Ang Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2311.00314v2 Announce Type: replace Abstract: Topic modeling has emerged as a valuable tool for discovering patterns and topics within large collections of documents. However, when cross-analysis involves multiple parties, data privacy becomes a critical concern. Federated topic modeling has b...

📖 Read original article


315. Randomized Confidence Bounds for Stochastic Partial Monitoring ​

Author: Maxime Heuillet, Ola Ahmad, Audrey Durand
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2402.05002v3 Announce Type: replace Abstract: The partial monitoring (PM) framework provides a theoretical formulation of sequential learning problems with incomplete feedback. On each round, a learning agent plays an action while the environment simultaneously chooses an outcome. The agent th...

📖 Read original article


316. Distributionally Robust Reinforcement Learning with Interactive Data Collection: Fundamental Hardness and Near-Optimal Algorithms ​

Author: Miao Lu, Han Zhong, Tong Zhang, Jose Blanchet
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2404.03578v3 Announce Type: replace Abstract: The sim-to-real gap, which represents the disparity between training and testing environments, poses a significant challenge in reinforcement learning (RL). A promising approach to addressing this challenge is distributionally robust RL, often fram...

📖 Read original article


317. Neural Active Learning Meets the Partial Monitoring Framework ​

Author: Maxime Heuillet, Ola Ahmad, Audrey Durand
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2405.08921v2 Announce Type: replace Abstract: We focus on the online-based active learning (OAL) setting where an agent operates over a stream of observations and trades-off between the costly acquisition of information (labelled observations) and the cost of prediction errors. We propose a no...

📖 Read original article


318. Constrained Reinforcement Learning for Safe Heat Pump Control ​

Author: Baohe Zhang, Lilli Frison, Thomas Brox, Joschka B"odecker
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY

arXiv:2409.19716v2 Announce Type: replace Abstract: Constrained Reinforcement Learning (RL) has emerged as a significant research area within RL, where integrating constraints with rewards is crucial for enhancing safety and performance across diverse control tasks. In the context of heating systems...

📖 Read original article


319. Training on Irrelevant States Implies Data Augmentation: Generalization in Contextual MDPs ​

Author: Max Weltevrede, Caroline Horsch, Matthijs T. J. Spaan, Wendelin B"ohmer
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2410.03565v4 Announce Type: replace Abstract: In the zero-shot policy transfer (ZSPT) setting for contextual Markov decision processes (CMDP), agents train on a fixed, finite set of contexts and must generalize to new ones. Recent work has demonstrated that training on additional states, even ...

📖 Read original article


320. Disentangled Unsupervised Skill Discovery for Efficient Hierarchical Reinforcement Learning ​

Author: Jiaheng Hu, Zizhao Wang, Peter Stone, Roberto Mart'in-Mart'in
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2410.11251v2 Announce Type: replace Abstract: A hallmark of intelligent agents is the ability to learn reusable skills purely from unsupervised interaction with the environment. However, existing unsupervised skill discovery methods often learn entangled skills where one skill variable simulta...

📖 Read original article


321. Device-Cloud Collaborative LLM Inference with Multi-Modal, Multi-Task, Multi-Turn Conversations ​

Author: Liangqi Yuan, Dong-Jun Han, Shiqiang Wang, Christopher G. Brinton
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2502.11007v5 Announce Type: replace Abstract: Compared to traditional machine learning models, recent large language models (LLMs) can exhibit multi-task-solving capabilities through multi-modal data sources and multi-turn conversations. These unique characteristics of LLMs, together with thei...

📖 Read original article


322. Training Diagonal Linear Networks with Stochastic Sharpness-Aware Minimization ​

Author: Gabriel Clara, Sophie Langer, Johannes Schmidt-Hieber
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH

arXiv:2503.11891v2 Announce Type: replace Abstract: We analyze the landscape and training dynamics of diagonal linear networks in a linear regression task, with the network parameters being perturbed by isotropic normal noise during training. The addition of such noise may be interpreted as a stocha...

📖 Read original article


323. Meta-Dependence in Conditional Independence Testing ​

Author: Bijan Mazaheri, Jiaqi Zhang, Caroline Uhler
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, stat.ML

arXiv:2504.12594v2 Announce Type: replace Abstract: Conditional independence testing is a critical component of feature screening, invariant statistical models, and causal discovery. Many of these algorithms rely on the sequential application of conditional independence tests, and their stability hi...

📖 Read original article


324. Riemannian Denoising Diffusion Probabilistic Models ​

Author: Zichen Liu, Wei Zhang, Christof Sch"utte, Tiejun Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2505.04338v3 Announce Type: replace Abstract: We propose Riemannian Denoising Diffusion Probabilistic Models (RDDPMs) for learning distributions on submanifolds of Euclidean space that are level sets of functions, including most of the manifolds relevant to applications. Existing methods for g...

📖 Read original article


325. A Learning-Based Ansatz Satisfying Boundary Conditions in Variational Problems ​

Author: Rafael Florencio, Julio Guerrero
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2505.12430v2 Announce Type: replace Abstract: Recently, innovative adaptations of the Ritz method incorporating deep learning have been developed, known as the Deep Ritz Method. This approach employs a neural network as the trial function for variational problems. However, the neural network d...

📖 Read original article


326. Efficient Q-Learning and Actor-Critic Methods for Robust Average-Reward Reinforcement Learning ​

Author: Yang Xu, Swetha Ganesh, Vaneet Aggarwal
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2506.07040v4 Announce Type: replace Abstract: We study model-free methods for distributionally robust infinite-horizon average-reward Markov decision processes (MDPs). We present non-asymptotic convergence analyses of Q-learning and actor-critic algorithms for robust average-reward MDPs under ...

📖 Read original article


327. Adaptive Reinforcement Learning for Unobservable Random Delays ​

Author: John Wikman, Alexandre Proutiere, David Broman
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO

arXiv:2506.14411v2 Announce Type: replace Abstract: In standard reinforcement learning (RL) settings, the interaction between the agent and the environment is typically modeled as a Markov decision process (MDP), which assumes that the agent observes the system state instantaneously, selects an acti...

📖 Read original article


328. On the Necessity of Output Distribution Reweighting for Effective Class Unlearning ​

Author: Ali Ebrahimpour-Boroojeny, Yian Wang, Hari Sundaram
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2506.20893v5 Announce Type: replace Abstract: In this paper, we reveal a significant shortcoming in class unlearning evaluations: overlooking the underlying class geometry can cause information leakage about the forgotten class. We further propose a simple unlearning strategy to mitigate this ...

📖 Read original article


329. Memory Savings at What Cost? A Study of Alternatives to Backpropagation ​

Author: Kunjal Panchal, Sunav Choudhary, Yuriy Brun, Hui Guan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2506.21833v2 Announce Type: replace Abstract: Forward-mode automatic differentiation (FmAD) and zero-order (ZO) optimization are increasingly proposed as memory-efficient, backpropagation-free alternatives for large language model (LLM) fine-tuning, yet their benefits are typically evaluated o...

📖 Read original article


330. Hyper-modal Imputation Diffusion Embedding with Dual-Distillation for Federated Multimodal Knowledge Graph Completion ​

Author: Ying Zhang, Yu Zhao, Xuhui Sui, Baohang Zhou, Xiangrui Cai, Li Shen, Xiaojie Yuan, Dacheng Tao
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.MM

arXiv:2506.22036v2 Announce Type: replace Abstract: With the increasing multimodal knowledge privatization requirements, multimodal knowledge graphs in different institutes are usually decentralized, lacking of effective collaboration system with both stronger reasoning ability and transmission safe...

📖 Read original article


Author: Xiaoya Li, Albert Wang, Guoyin Wang, Chris Shum, Jiwei Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.DB

arXiv:2508.02091v3 Announce Type: replace Abstract: Approximate nearest-neighbor search (ANNS) algorithms have become increasingly critical for recent AI applications, particularly in retrieval-augmented generation (RAG) and agent-based LLM applications. In this paper, we present CRINN, a new paradi...

📖 Read original article


332. Beyond Na\"ive Prompting: Strategies for Improved Context-aided Forecasting with LLMs ​

Author: Arjun Ashok, Andrew Robert Williams, Vincent Zhihao Zheng, Irina Rish, Nicolas Chapados, 'Etienne Marcotte, Valentina Zantedeschi, Alexandre Drouin
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2508.09904v3 Announce Type: replace Abstract: Real-world forecasting requires models to integrate not only historical data but also relevant contextual information provided in textual form. While large language models (LLMs) show promise for context-aided forecasting, critical challenges remai...

📖 Read original article


333. Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts ​

Author: Maxime Heuillet, Yufei Cui, Boxing Chen, Audrey Durand, Prasanna Parthasarathi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2508.10123v3 Announce Type: replace Abstract: Advanced reasoning in LLMs on challenging domains like mathematical reasoning can be tackled using verifiable rewards based reinforced fine-tuning (ReFT). In standard ReFT frameworks, a behavior model generates multiple completions with answers per...

📖 Read original article


334. A Data-Driven Interpolation Method on Smooth Manifolds via Diffusion Processes and Voronoi Tessellations ​

Author: Alvaro Almeida Gomez
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2509.03758v5 Announce Type: replace Abstract: We propose a data-driven interpolation framework for reconstructing real-valued functions on smooth manifolds from scattered pointwise observations. The method combines a Gaussian Nadaraya--Watson kernel interpolant with a Voronoi-adaptive bandwidt...

📖 Read original article


335. Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints ​

Author: Francesco Emanuele Stradi, Eleonora Fidelia Chiefari, Matteo Castiglioni, Alberto Marchesi, Nicola Gatti
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.20114v3 Announce Type: replace Abstract: We study \emph{online episodic Constrained Markov Decision Processes} (CMDPs) under both stochastic and adversarial constraints. We provide a novel algorithm whose guarantees greatly improve those of the state-of-the-art best-of-both-worlds algorit...

📖 Read original article


336. Graph Optimization Foundation Model: Tokenizing Graph via A Language-Model Paradigm ​

Author: Yunhao Liang, Pujun Zhang, Yuan Qu, Jingyuan Yang, Shaochong Lin, Zuo-jun Max Shen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2509.24256v2 Announce Type: replace Abstract: The pretrain-transfer paradigm, which underpins the success of large language models (LLMs), has demonstrated the immense power of creating foundation models that learn generalizable representations from vast datasets. However, extending this parad...

📖 Read original article


337. Unveiling the Mechanisms of Multi-Hop Reasoning in Transformers via Identity Bridge ​

Author: Pengxiao Lin, Zheng-An Chen, Zhi-Qin John Xu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2509.24653v2 Announce Type: replace Abstract: Large Language Models (LLMs) excel at multi-hop reasoning in distribution, yet fail on unseen compositions, a phenomenon known as the curse of two-hop reasoning. In this work, we argue that this phenomenon can be attributed to a missing supervision...

📖 Read original article


338. Enhancing Reasoning for Diffusion LLMs via Distribution Matching Policy Optimization ​

Author: Yuchen Zhu, Wei Guo, Jaemoo Choi, Petr Molodyk, Bo Yuan, Molei Tao, Yongxin Chen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.08233v3 Announce Type: replace Abstract: Diffusion large language models (dLLMs) are promising alternatives to autoregressive large language models (AR-LLMs), as they potentially allow higher inference throughput. Reinforcement learning (RL) is crucial to enabling dLLMs to achieve perform...

📖 Read original article


339. OpenEM: Large-scale multi-structural 3D datasets for electromagnetic methods ​

Author: Shuang Wang, Xuben Wang, Fei Deng, Peifan Jiang, Jian Chen, Gianluca Fiandaca
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.21859v3 Announce Type: replace Abstract: Electromagnetic (EM) methods, owing to their efficiency and non-invasive nature, have become one of the most widely used techniques in geological exploration. Nevertheless, data processing for these methods remains highly time-consuming and labor-i...

📖 Read original article


340. CANDI: Hybrid Discrete-Continuous Diffusion Models ​

Author: Patrick Pynadath, Jiaxin Shi, Ruqi Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2510.22510v3 Announce Type: replace Abstract: While continuous diffusion has shown remarkable success in continuous domains such as image generation, its direct application to discrete data has underperformed pure discrete formulations. To understand this gap, we introduce token identifiabilit...

📖 Read original article


341. Controllably Efficient Language Models ​

Author: Jatin Prakash, Aahlad Puli, Rajesh Ranganath
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.05313v2 Announce Type: replace Abstract: The substantial inference costs of attention in transformers motivated the development of efficient sequence mixers: namely sparse and sliding window attention, convolutions and linear attention. Although these approaches result in impressive reduc...

📖 Read original article


342. Enabling Agents to Communicate Entirely in Latent Space ​

Author: Zhuoyun Du, Runze Wang, Huiyu Bai, Zouying Cao, Xiaoyong Zhu, Yu Cheng, Bo Zheng, Wei Chen, Haochao Ying
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA

arXiv:2511.09149v5 Announce Type: replace Abstract: While natural language is the de facto communication medium for LLM-based agents, it presents a fundamental constraint. The process of downsampling rich, internal latent states into discrete tokens inherently limits the depth and nuance of informat...

📖 Read original article


343. Enhancing Adversarial Transferability through Block Stretch and Shrink ​

Author: Quan Liu, Feng Ye, Chenhao Lu, Shuming Zhen, Guanliang Huang, Lunzhe Chen, Xudong Ke
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2511.17688v2 Announce Type: replace Abstract: Input transformation-based attacks improve adversarial transferability by aggregating gradients over transformed inputs. Existing analyses mainly explain their efficacy from image diversity, semantic preservation, attention variance or hypothesis s...

📖 Read original article


344. CUDA-L2: Surpassing cuBLAS Performance for Matrix Multiplication through Reinforcement Learning ​

Author: Songqiao Su, Xiaoya Li, Albert Wang, Guoyin Wang, Jiwei Li, Chris Shum
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2512.02551v3 Announce Type: replace Abstract: In this paper, we propose CUDA-L2, a system that combines large language models (LLMs) and reinforcement learning (RL) to automatically optimize Half-precision General Matrix Multiply (HGEMM) CUDA kernels. Using CUDA execution speed as the RL rewar...

📖 Read original article


345. From Confounding to Learning: Dynamic Service Fee Pricing on Third-Party Platforms ​

Author: Rui Ai, David Simchi-Levi, Feng Zhu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.22749v2 Announce Type: replace Abstract: We study the pricing behavior of third-party platforms facing strategic agents. Assuming the platform is a revenue maximizer, it observes market features that generally affect demand. Since only transacted quantities and prices can be observed, thi...

📖 Read original article


346. Stable On-Policy Distillation through Adaptive Target Reformulation ​

Author: Ijun Jang, Jewon Yeom, Juan Yeo, Hyunggyu Lim, Taesup Kim
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.07155v3 Announce Type: replace Abstract: Knowledge distillation (KD) is a widely adopted technique for transferring knowledge from large language models to smaller student models; however, conventional supervised KD often suffers from a distribution mismatch between training and inference...

📖 Read original article


347. BalDRO: A Distributionally Robust Optimization based Framework for Large Language Model Unlearning ​

Author: Pengyang Shao, Naixin Zhai, Lei Chen, Yonghui Yang, Fengbin Zhu, Xun Yang, Meng Wang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.09172v3 Announce Type: replace Abstract: As Large Language Models (LLMs) increasingly shape online content, removing targeted information from well-trained LLMs (also known as LLM unlearning) has become critical for web governance. A key challenge lies in sample-wise imbalance within the ...

📖 Read original article


348. TimeSAE: Causal Sparse Decoding for Faithful Explanations of Black-Box Time Series Models ​

Author: Khalid Oublal, Quentin Bouniot, Qi Gan, Stephan Cl'emen\c{c}on, Zeynep Akata
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.09776v2 Announce Type: replace Abstract: As black box models and pretrained models gain traction in time series applications, understanding and explaining their predictions becomes increasingly vital, especially in high-stakes domains where interpretability and trust are essential. Howeve...

📖 Read original article


349. Multimodal rumor detection enhanced by external evidence and forgery features ​

Author: Han Li, Hua Sun
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.14954v3 Announce Type: replace Abstract: Social media increasingly disseminates information through mixed image text posts, but rumors often exploit subtle inconsistencies and forged content, making detection based solely on post content difficult. Deep semantic mismatch rumors, which sup...

📖 Read original article


350. HyperNet-Adaptation for Diffusion-Based Test Case Generation ​

Author: Oliver Wei{\ss}l, Vincenzo Riccio, Severin Kacianka, Andrea Stocco
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.SE

arXiv:2601.15041v2 Announce Type: replace Abstract: The increasing deployment of deep learning systems requires systematic evaluation of their reliability in real-world scenarios. Traditional gradient-based adversarial attacks introduce small perturbations that rarely correspond to realistic failure...

📖 Read original article


351. SFO: Learning PDE Operators via Spectral Filtering ​

Author: Noam Koren, Rafael Moschopoulos, Kira Radinsky, Elad Hazan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.17090v2 Announce Type: replace Abstract: Partial differential equations (PDEs) govern complex systems, yet neural operators often struggle to efficiently capture the long-range, nonlocal interactions inherent in their solution maps. We introduce Spectral Filtering Operator (SFO), a neural...

📖 Read original article


352. Explainability Methods for Hardware Trojan Detection: A Systematic Comparison ​

Author: Paul Whitten, Francis Wolff, Chris Papachristou
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.18696v5 Announce Type: replace Abstract: Hardware trojans are malicious circuits which compromise the functionality and security of an integrated circuit (IC). These circuits are manufactured directly into the silicon and cannot be fixed by security patches like software. The solution wou...

📖 Read original article


353. Rethinking Zero-Shot Time Series Classification: From Task-specific Classifiers to In-Context Inference ​

Author: Juntao Fang, Shifeng Xie, Shengbin Nie, Yuhui Ling, Yuming Liu, Zijian Li, Keli Zhang, Lujia Pan, Themis Palpanas, Ruichu Cai
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.00620v2 Announce Type: replace Abstract: The zero-shot evaluation of time series foundation models (TSFMs) for classification typically uses a frozen encoder followed by a task-specific classifier. However, this practice violates the training-free premise of zero-shot deployment and intro...

📖 Read original article


354. CoGenCast: A Coupled Autoregressive-Flow Generative Framework for Time Series Forecasting ​

Author: Mingyue Cheng, Yaguo Liu, Daoyu Wang, Xiaoyu Tao, Qi Liu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.03564v2 Announce Type: replace Abstract: Time series forecasting can be viewed as a generative problem that requires both semantic understanding over contextual conditions and stochastic modeling of continuous temporal dynamics. Existing approaches typically rely on either autoregressive ...

📖 Read original article


355. Disentangling Intrinsic Importance from Emergent Structure in Multi-Expert Orchestration ​

Author: Sudipto Ghosh, Sujoy Nath, Sunny Manchanda, Tanmoy Chakraborty
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA

arXiv:2602.04291v3 Announce Type: replace Abstract: Multi-expert systems, where multiple Large Language Models (LLMs) collaborate to solve complex tasks, are increasingly adopted for high-performance reasoning and generation. However, the orchestration policies governing expert interaction and seque...

📖 Read original article


356. Separation-Utility Pareto Frontier: An Information-Theoretic Characterization ​

Author: Shizhou Xu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2602.04408v3 Announce Type: replace Abstract: We study the Pareto frontier (optimal trade-off) between utility and separation, a fairness criterion requiring predictive independence from sensitive attributes conditional on the true outcome. Through an information-theoretic lens, we prove a cha...

📖 Read original article


357. Probabilistic Wind Power Forecasting with Tree-Based Machine Learning and Weather Ensembles ​

Author: Max Bruninx, Diederik van Binsbergen, Timothy Verstraeten, Ann Now'e, Jan Helsen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.13010v2 Announce Type: replace Abstract: Accurate production forecasts are essential for the integration of renewable energy sources into the power grid. This paper illustrates how to obtain probabilistic forecasts of wind power generation using gradient boosting trees and an ensemble of ...

📖 Read original article


358. SynthSAEBench: Evaluating Sparse Autoencoders on Scalable Realistic Synthetic Data ​

Author: David Chanin, Adri`a Garriga-Alonso
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.14687v2 Announce Type: replace Abstract: Improving Sparse Autoencoders (SAEs) requires benchmarks that can precisely validate architectural innovations. Current LLM-based SAE benchmarks are too noisy to differentiate architectural improvements, while commonly used synthetic-data experimen...

📖 Read original article


359. BRIDGE: Bridging Reasoning In Distillation Gap Elimination via Structure-Aware Masking ​

Author: Bowen Yu, Sheng Zhang, Binhao Wang, Yi Wen, Jingtong Gao, Bowen Liu, Zimo Zhao, Shanshan Ye, Wanyu Wang, Maolin Wang, Xiangyu Zhao
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17686v5 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning has significantly improved LLMs' mathematical problem-solving capabilities, but distilling such capabilities into smaller models remains challenging due to the capacity mismatch between verbose teachers and compact ...

📖 Read original article


360. Turbo Connection: Reasoning as Information Flow from Higher to Lower Layers ​

Author: Mohan Tang, Sidi Lu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17993v2 Announce Type: replace Abstract: Complex problems, whether in math, logic, or planning, are solved by humans through a sequence of steps where the result of one step informs the next. In this work, we adopt the perspective that the reasoning power of Transformers is fundamentally ...

📖 Read original article


361. Dynamics of Learning under User Choice: Overspecialization and Peer-Model Probing ​

Author: Adhyyan Narang, Sarah Dean, Lillian J Ratliff, Maryam Fazel
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.MA

arXiv:2602.23565v2 Announce Type: replace Abstract: In many economically relevant contexts where machine learning is deployed, multiple platforms obtain data from the same pool of users, each of whom selects the platform that best serves them. Prior work in this setting focuses exclusively on the "l...

📖 Read original article


362. Autocorrelation effects in a stochastic-process model for solving two-armed bandit problems ​

Author: Tomoki Yamagami, Mikio Hasegawa, Takatomo Mihana, Ryoichi Horisaki, Atsushi Uchida
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.ET, math.PR, physics.optics

arXiv:2603.05559v2 Announce Type: replace Abstract: Decision makers exploiting photonic chaotic dynamics obtained by semiconductor lasers provide an ultrafast approach to solving multi-armed bandit problems by using a temporal optical signal as the driving source for sequential decisions. In such sy...

📖 Read original article


363. First-Order Softmax Weighted Switching Gradient Method for Distributed Stochastic Minimax Optimization with Stochastic Constraints ​

Author: Zhankun Luo, Antesh Upadhyay, Sang Bin Moon, Abolfazl Hashemi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2603.05774v2 Announce Type: replace Abstract: This paper addresses the distributed stochastic minimax optimization problem subject to stochastic constraints. We propose a novel first-order Softmax-Weighted Switching Gradient method tailored for federated learning. Under full client participati...

📖 Read original article


364. Local Message-Passing for Discrete Graph Generation ​

Author: Jay Revolinsky, Harry Shomer, Jiliang Tang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.08825v2 Announce Type: replace Abstract: Discrete graph generation has emerged as a powerful paradigm for modeling graph-structured data, yet state of the art models often rely on Graph Transformers or higher order architectures. We revisit this design assumption by introducing GenGNN, a ...

📖 Read original article


365. ECoLAD: Selecting Anomaly Detectors for Automotive Deployment via Compute-Reduction Evaluation ​

Author: Kadir-Kaan "Ozer, Ren'e Ebeling, Markus Enzweiler
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.10926v2 Announce Type: replace Abstract: Automotive anomaly detectors are often selected from accuracy only benchmarks on workstation class hardware, whereas in-vehicle monitoring requires predictable scoring latency under limited CPU parallelism. This mismatch can make methods that appea...

📖 Read original article


366. Simple Recipe Works: Vision-Language-Action Models are Natural Continual Learners with Reinforcement Learning ​

Author: Jiaheng Hu, Jay Shim, Chen Tang, Yoonchang Sung, Bo Liu, Peter Stone, Roberto Martin-Martin
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2603.11653v3 Announce Type: replace Abstract: Continual Reinforcement Learning (CRL) for Vision-Language-Action (VLA) models is a promising direction toward self-improving embodied agents that can adapt in openended, evolving environments. However, conventional wisdom from continual learning s...

📖 Read original article


367. The Cost of Reasoning: Chain-of-Thought Induces Overconfidence in Vision-Language Models ​

Author: Robert Welch, Emir Konuk, Kevin Smith
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.16728v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly deployed in high-stakes settings where reliable uncertainty quantification (UQ) is as important as predictive accuracy. Extended reasoning via chain-of-thought (CoT) prompting or reasoning-trained mode...

📖 Read original article


368. How Can Machine Learning Emulators Best Support Climate Science? ​

Author: Luca Schmidt, Nina Effenberger, Vitus Benson, Philine L. Bommer, Robert Brunstein, Mikel N. Legasa, Maxim Samarin, Maybritt Schillinger
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.AP, stat.ML

arXiv:2603.22320v3 Announce Type: replace Abstract: For decades, physics-based climate models have been used to provide insights for climate decision-making. Their application is, however, constrained by significant computational and technical demands. Machine learning (ML) emulators offer a way to ...

📖 Read original article


369. Critical Damping as a Momentum Schedule: Multi-Seed Validation, a Hybrid Recipe, and an Exhaustive Negative Result on Surgical Layer Selection ​

Author: Ivan Pasichnyk
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.28921v3 Announce Type: replace Abstract: The critical damping condition of the damped harmonic oscillator model of SGD with momentum (Qian, 1999) yields a momentum schedule with no tuned hyperparameters: mu(t) = 1 - 2*sqrt(alpha(t)). Across five seeds on ResNet-18/CIFAR-10 (200-epoch cosi...

📖 Read original article


370. Neural Collapse Dynamics: Depth, Activation, Regularisation, and Feature Norm Threshold ​

Author: Anamika Paul Rupa
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.00230v3 Announce Type: replace Abstract: Neural collapse (NC) -- the convergence of penultimate-layer features to a simplex equiangular tight frame -- is well understood at equilibrium, but the dynamics governing its onset remain poorly characterised. We identify a simple and predictive r...

📖 Read original article


371. ACES: Who Tests the Tests? Leave-One-Out AUC Consistency for Code Generation ​

Author: Hui Sun, Yun-Ji Zhang, Zheng Xie, Ren-Biao Liu, Yali Du, Xin-Ye Li, Ming Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.03922v2 Announce Type: replace Abstract: Selecting LLM-generated code candidates using LLM-generated tests is challenging because the tests themselves may be incorrect. Existing methods either treat all tests equally or rely on ad-hoc heuristics to filter unreliable tests. Yet determining...

📖 Read original article


372. Observable Performance Does Not Fully Reflect Adaptive System Organization: A Multi-Level Analysis of Gait Dynamics Under Occlusal Constraint ​

Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC

arXiv:2605.00778v2 Announce Type: replace Abstract: In biomechanical systems, observable performance is often used as a proxy for underlying organization, although similar outputs may arise from different adaptive configurations. This study considers the vertical dimension of occlusion (VDO) as a co...

📖 Read original article


373. Congestion-Aware Dynamic Axonal Delay for Spiking Neural Networks ​

Author: Dewei Bai, Hongxiang Peng, Yunyun Zeng, Ziyu Zhang, Hong Qu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.01291v3 Announce Type: replace Abstract: Spiking Neural Networks (SNNs) are widely regarded as an energy-efficient paradigm for modeling and processing temporal and event-driven information. Incorporating delays in SNNs has been proven to be an effective mechanism for improving spike alig...

📖 Read original article


374. Retrieval with Multiple Query Vectors through Anomalous Pattern Detection ​

Author: Allassan Tchangmena A Nken, Baimam Boukar Jean Jacques, Miriam Rateike, Celia Cintas, Skyler Speakman
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.01965v2 Announce Type: replace Abstract: A classical vector retrieval problem typically considers a \emph{single} query embedding vector as input and retrieves the most similar embedding vectors from a vector database. However, complex reasoning and retrieval tasks frequently require \emp...

📖 Read original article


375. Attribution-Guided Continual Learning for Large Language Models ​

Author: Yazheng Liu, Yuxuan Wan, Rui Xu, Xi Zhang, Sihong Xie, Hui Xiong
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.05285v2 Announce Type: replace Abstract: Large language models (LLMs) often suffer from catastrophic forgetting in continual learning: after learning new tasks sequentially, they perform worse on earlier tasks. Existing methods mitigate catastrophic forgetting by data replay, parameter fr...

📖 Read original article


376. COSMOS: Model-Agnostic Personalized Federated Learning with Clustered Server Models and Pseudo-Label-Only Communication ​

Author: Ben Rachmut, Luise Ge, William Yeoh, Ning Zhang, Yevgeniy Vorobeychik
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.11165v3 Announce Type: replace Abstract: Federated learning (FL) in heterogeneous environments remains challenging because client models often differ in both architecture and data distribution. While recent approaches attempt to address this challenge through client clustering and knowled...

📖 Read original article


377. CTFusion: A CTF-based Benchmark for LLM Agent Evaluation ​

Author: Dongjun Lee, Ga-eun Bae, Insu Yun
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2605.11504v2 Announce Type: replace Abstract: Recent advances in Large Language Models (LLMs) have enabled agentic systems for complex, multi-step tasks; cybersecurity is emerging as a prominent application. To evaluate such agents, researchers widely adopt Capture The Flag (CTF) benchmarks. H...

📖 Read original article


378. Selective Safety Steering via Value-Filtered Decoding ​

Author: Bat-Sheva Einbinder, Hen Davidov, Yee Whye Teh, Yarin Gal, Yaniv Romano
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.14746v2 Announce Type: replace Abstract: While large language models (LLMs) are trained to align with human values, their generations may still violate safety constraints. A growing line of work addresses this problem by modifying the model's sampling policy at decoding time using a safet...

📖 Read original article


379. Causal Foundation Models with Continuous Treatments ​

Author: Christopher Stith, Medha Barath, Vahid Balazadeh, Jesse C. Cresswell, Rahul G. Krishnan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.15133v2 Announce Type: replace Abstract: Causal inference, estimating causal effects from observational data, is a fundamental tool in many disciplines. Of particular importance across a variety of domains is the continuous treatment setting, where the variable of intervention has a conti...

📖 Read original article


380. From Observed Viability to Internal Predictive Approximation: A Single-Subject Latent-Space Analysis of Gait Dynamics Under Occlusal Constraint ​

Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC

arXiv:2605.15862v2 Announce Type: replace Abstract: Understanding adaptive biomechanical systems requires distinguishing observable performance, static multivariate representation, longitudinal displacement, and internal approximation of observed change. This study introduces Level 5, which examines...

📖 Read original article


381. Prune, Update and Trim: Robust Structured Pruning for Large Language Models ​

Author: Diego Coello de Portugal Mecke, Tom Hanika, Lars Schmidt-Thieme
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.18331v2 Announce Type: replace Abstract: Large Language Models (LLMs) have experienced significant growth and development in recent years. However, performing inference on LLMs remains costly, especially for long-context inference or in resource-constrained devices. This motivates the dev...

📖 Read original article


382. Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs ​

Author: Qihong Yang, Qiaolin He
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.31027v4 Announce Type: replace Abstract: Solving high-frequency partial differential equations (PDEs) with neural networks is notoriously difficult due to the spectral bias of conventional architectures. We propose the Multi-Scale Separable Fourier Neural Network (MS-SFNN), a framework de...

📖 Read original article


383. Student Capacity Moderates Knowledge Distillation Effectiveness: A Systematic Study Across ResNet Teacher-Student Pairs on CIFAR-10 ​

Author: Umut Onur Yasar
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2605.31191v2 Announce Type: replace Abstract: We investigate how teacher-student capacity relationships modulate knowledge distillation (KD) effectiveness in ResNet-based image classification on CIFAR-10. Across four teacher-student pairs (R50->R18, R34->R18, R50->R34, and R101->R34) we compar...

📖 Read original article


384. Revisiting Neural Processes via Fourier Transform and Volterra Series ​

Author: Peiman Mohseni, Nick Duffield, Raymond K. W. Wong
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2606.01172v3 Announce Type: replace Abstract: Modeling unknown latent functions from finite, irregularly sampled measurements is a recurring challenge across science and engineering. Neural processes (NPs), a family of probabilistic functional models, are promising solutions -- especially when...

📖 Read original article


385. From Performance to Representational Adequacy: A Representational Bootstrap Framework for Adaptive Biological Systems ​

Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.01374v3 Announce Type: replace Abstract: Observable performance is commonly used to characterize biological systems, yet aggregated outputs may remain insufficient for uniquely resolving observational conditions, and richer multivariate representations may retain substantial ambiguity. Th...

📖 Read original article


386. IntraShuffler: A Privacy Preserving Framework for Heterogeneous DP Federated Learning ​

Author: Farhin Farhad Riya, Olivera Kotevska, Jinyuan Stella Sun
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.DC

arXiv:2606.02563v2 Announce Type: replace Abstract: Heterogeneous Differential Privacy (HDP) in Federated Learning (FL) allows clients to select individual privacy budgets ($\varepsilon_i$) according to institutional policies and data sensitivity. In practice, many HDP-FL systems employ $\varepsilon...

📖 Read original article


387. Online KL-Regularized Reinforcement Learning with Function Approximation under Misspecification ​

Author: Haoyang Hong, Zichen Wang, Quanquan Gu, Huazheng Wang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.06053v2 Announce Type: replace Abstract: We study KL-regularized contextual bandits and episodic reinforcement learning (RL) under general function approximation with model misspecification. Existing guarantees rely on realizability and therefore do not extend to misspecified models, wher...

📖 Read original article


388. Bootstrap Theory of Representational Emergence: Explanatory Insufficiency as a Driver of Representation Learning and World Models ​

Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.07303v2 Announce Type: replace Abstract: Representation learning is central to modern machine learning, but most research examines how representations are optimized after a framework has been selected. Less attention is given to when a new representational level becomes necessary. This ar...

📖 Read original article


389. The Injection Paradox: Brand-Level Suppression in Safety-Trained LLM Recommendations via RAG Context Injection ​

Author: Hyunseok Paeng
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CR

arXiv:2606.09204v2 Announce Type: replace Abstract: We present a reproducible failure mode of safety training in RAG-based LLM recommendation, the Injection Paradox, in which prompt injections embedded in retrieved documents backfire against the attacker, suppressing the target brand below the injec...

📖 Read original article


390. Learning Dynamics Reveal a Hierarchy of Weight-Induced Layerwise Gram Metrics ​

Author: Claudio Nordio
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn

arXiv:2606.09744v4 Announce Type: replace Abstract: We study feed-forward ReLU networks with fixed readout and quadratic loss, and rewrite gradient descent as a collective dynamics of activation fields and conjugate fields on the training set. Working to first order in the learning rate inside a fix...

📖 Read original article


391. Detecting Explanatory Insufficiency in Learned Representations: A Framework for Representational Vigilance ​

Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.13172v2 Announce Type: replace Abstract: Learned representations are central to modern machine learning and are commonly evaluated through predictive performance, robustness, uncertainty estimation, and generalization. However, a representation may remain operationally successful while fa...

📖 Read original article


392. Gefen: Optimized Stochastic Optimizer ​

Author: Nadav Benedek, Tomer Koren, Ohad Fried
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CV

arXiv:2606.13894v2 Announce Type: replace Abstract: AdamW is a default optimizer for modern deep learning, but its first and second moment states add roughly two parameter-sized buffers to training memory, increasing the already substantial cost of large-scale pretraining. We propose Gefen, a memory...

📖 Read original article


393. FAST: A Framework for Aligned Sampling and Training in Parallel Reinforcement Learning for Autonomous Driving ​

Author: Bonan Wang, Letian Tao, Bin Shuai, Jiaxin Gao, Wenxin Zhao, Wei Xiong, Kehua Sheng, Bo Zhang, Yang Guan, Shengbo Eben Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.21587v2 Announce Type: replace Abstract: Deep reinforcement learning is pivotal for closed-loop autonomous driving yet remains constrained by severe bottlenecks in sampling efficiency. Standard parallel sampling mitigates this but suffers from the straggler effect, where the premature ter...

📖 Read original article


394. The Geometry of Saturation: Effective Rank Predicts When Labels Stop Helping in Few-Shot Classification ​

Author: Arnav Gupta
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.24903v2 Announce Type: replace Abstract: Few-shot label acquisition lacks a label-free signal for when additional labels cease to improve accuracy: existing stopping criteria either require a held-out validation set (violating the few-shot premise) or rely on theoretically ungrounded heur...

📖 Read original article


395. Multi-Agent Routing as Set-Valued Prediction: A WildChat Benchmark and Cost-Aware Evaluation ​

Author: Ananto Nayan Bala, Faisal Muhammad Shah
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR, cs.MA

arXiv:2606.28925v2 Announce Type: replace Abstract: Tool and agent routing from natural-language prompts is naturally a set-valued prediction problem: a single query may require multiple agents, while over-selection increases execution cost. The benchmark introduced here is derived from WildChat and...

📖 Read original article


396. Bridging the Gap Between Latent and Explicit Reasoning with Looped Transformers ​

Author: Ying Fan, Anej Svete, Kangwook Lee
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2606.31779v2 Announce Type: replace Abstract: Language models typically reason via explicit chain-of-thought (CoT), generating intermediate steps token-by-token. Latent CoT offers an alternative: it performs multi-step reasoning in the model's hidden states, replacing decoded tokens with conti...

📖 Read original article


397. DemoPSD: Disagreement-Modulated Policy Self-Distillation ​

Author: Yunhe Li, Hao Shi, Wenhao Liu, Mengzhe Ruan, Hanxu Hou, Zhongxiang Dai, Shuang Qiu, Linqi Song
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.02502v3 Announce Type: replace Abstract: On-policy self-distillation (OPSD) has emerged as a practical method for training large language models (LLMs) to reason, where a single model acts as both the teacher and the student with different levels of information access. However, recent stu...

📖 Read original article


398. A Structural Interpretation of GELU and Threshold-Transmission Activations via the First-Order Loss Function ​

Author: Roberto Rossi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2607.03664v2 Announce Type: replace Abstract: The Gaussian Error Linear Unit is usually motivated as the expected output of an input-dependent Bernoulli gate. This work gives an alternative interpretation: GELU is the expected output of a hard linear gate with a Gaussian random threshold. This...

📖 Read original article


399. Directional Curvature from Armijo Backtracking: A Low-Cost Sharpness Probe and a Calibration-Free Learning-Rate Safeguard for Adam ​

Author: Ashmitha R, J"org Frochte
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.03998v2 Announce Type: replace Abstract: The local sharpness of the loss, the top Hessian eigenvalue $\lambda_1$, determines the largest stable gradient step, but measuring it normally requires Lanczos or Hessian-vector iterations. We observe that a single Armijo backtracking line search ...

📖 Read original article


400. RL Forgets! Towards Continual Policy Optimization ​

Author: Mao-Lin Luo, Zhe-Xu Wang, Zi-Hao Zhou, Bo Ye, Jian Zhao, Min-Ling Zhang, Tong Wei
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.04364v2 Announce Type: replace Abstract: Continual post-training is becoming a central paradigm for adapting vision-language models to evolving tasks. Recent work has increasingly favored reinforcement learning over supervised fine-tuning, driven by the belief that reinforcement learning ...

📖 Read original article


401. Minimum Block Width for Universal Approximation by Residual Neural Networks with Inner Width One ​

Author: Qi Zhou, Xuan Zhou, Xiao-Song Yang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.04597v2 Announce Type: replace Abstract: In this paper, we study the universal approximation property of residual neural networks, and obtain some new results. For input and output dimensions $d_x$ and $d_y$, and LeakyReLU, ReLU, ReLU-like activation functions, the upper and lower bounds ...

📖 Read original article


402. Layer-Parallel Inference Reduces Encrypted Nonlinear Depth in Transformers ​

Author: Ligong Han, Kai Xu, Hao Wang, Ruijiang Gao, Han Gao, Akash Srivastava
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2607.04819v2 Announce Type: replace Abstract: Fully homomorphic encryption (FHE) enables computation on encrypted data, but practical encrypted Transformer inference is bottlenecked by the sequential composition of many nonlinear blocks. We study whether Structured Newton Layer Parallelism (SN...

📖 Read original article


403. x-Prediction Is All You Need:Training-Free Accelerated Generation via Endpoint Decodability ​

Author: Xin Peng, Ang Gao
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.06114v2 Announce Type: replace Abstract: Diffusion and flow matching models generate high-quality samples, but their ODE samplers often need tens to hundreds of neural function evaluations (NFEs). This remains a practical challenge for released checkpoints, since many accelerators require...

📖 Read original article


404. Trees from Marginals: Autoregressive drafting with factorized priors ​

Author: Yuma Oda, Ryan Mathieu, Roman Knyazhitskiy, Artur Chakhvadze
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.06763v2 Announce Type: replace Abstract: Speculative decoding greatly increases the interactivity of autoregressive language models by trading off computation for extra tokens generated in a single forward pass. Factorized draft models are especially efficient because they predict future-...

📖 Read original article


405. Efficient Long-Horizon Learning for Learned Optimization ​

Author: Xiaolong Huang, Benjamin Th'erien, James Harrison, Eugene Belilovsky
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.06772v3 Announce Type: replace Abstract: Learned optimization aims to improve upon hand-designed optimizers (e.g., Adam and Muon) by meta-learning small neural network optimizers over a distribution of tasks. While recent work has greatly advanced the architectural design and inductive bi...

📖 Read original article


406. Eigenbasis-Independent Learnable Spectral Positional Encodings for Directed Graphs via Hermitian Block Krylov Subspaces ​

Author: Jiaqing Xie, Yuxin Wang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.07032v2 Announce Type: replace Abstract: Spectral positional encodings (PEs) for \emph{directed} graphs face two obstacles: magnetic Laplacians require an $O(n^3)$ Hermitian eigendecomposition per potential, and their complex eigenvectors are defined only up to unitary gauge, which prior ...

📖 Read original article


407. Causal Optimizer Interaction Calculus: Hidden Geometric Relaxation and Identifiable Interventions ​

Author: Zavier Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2607.07206v2 Announce Type: replace Abstract: Optimizer experiments observe responses to algorithmic configurations without uniquely revealing hidden mechanisms. We develop a causal optimizer interaction calculus that separates pathwise realization, Mobius decomposition, and experimental ident...

📖 Read original article


408. Robust Bayesian Decision Making under Adversarial Uncertainty ​

Author: Haripriya Harikumar, Sammie Katt, Yasir Zubayr Barlas, Samuel Kaski
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.08590v2 Announce Type: replace Abstract: Scientific experiments are often designed to maximize information gain, yet in many applications the primary objective is to support reliable downstream decision-making. Existing decision-aware experimental design and active learning methods typica...

📖 Read original article


409. LieBN: Batch Normalization over Lie Groups ​

Author: Ziheng Chen, Yue Song, Rui Wang, Xiao-Jun Wu, Nicu Sebe
Published: 7/14/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08783v2 Announce Type: replace Abstract: Manifold-valued measurements are prevalent in various machine learning tasks. Recent advances have extended Deep Neural Networks (DNNs) to operate on manifolds, accompanied by normalization techniques tailored to different geometries, collectively ...

📖 Read original article


410. Research on Intellectual Property Resource Profile and Evolution Law ​

Author: Yuhui Wang, Yingxia Shao, Ang Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.DL, cs.AI, cs.LG

arXiv:2204.06221v2 Announce Type: replace-cross Abstract: In the era of big data, intellectual property-oriented scientific and technological resources show the trend of large data scale, high information density, and low value density, which brings severe challenges to the effective use of intellec...

📖 Read original article


411. Interventions Against Machine-Assisted Statistical Discrimination ​

Author: John Y. Zhu
Published: 7/14/2026, 4:00:00 AM
Categories: econ.TH, cs.LG

arXiv:2310.04585v5 Announce Type: replace-cross Abstract: I study statistical discrimination driven by verifiable beliefs, such as those generated by machine learning, rather than by humans. When beliefs are verifiable, interventions against statistical discrimination can move beyond simple belief-f...

📖 Read original article


412. HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA ​

Author: Xinyue Chen, Pengyu Gao, Jiangjiang Song, Xiaoyang Tan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2402.01767v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) has rapidly advanced the language model field, particularly in question-answering (QA) systems. By integrating external documents during the response generation phase, RAG significantly enhances the accura...

📖 Read original article


413. Big data approach to Kazhdan-Lusztig polynomials ​

Author: Abel Lacabanne, Daniel Tubbenhauer, Pedro Vaz
Published: 7/14/2026, 4:00:00 AM
Categories: math.RT, cs.LG, math.CO

arXiv:2412.01283v3 Announce Type: replace-cross Abstract: We investigate the structure of Kazhdan-Lusztig polynomials of the symmetric group by leveraging computational approaches from big data, including exploratory and topological data analysis, applied to the polynomials for symmetric groups of u...

📖 Read original article


414. Low-dimensional adaptation of diffusion models: Convergence in total variation ​

Author: Jiadong Liang, Zhihan Huang, Yuxin Chen
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2501.12982v3 Announce Type: replace-cross Abstract: This paper investigates how diffusion generative models leverage (unknown) low-dimensional structure to accelerate sampling. Focusing on two mainstream samplers -- the denoising diffusion implicit model (DDIM) and the denoising diffusion prob...

📖 Read original article


415. Disentangling Feature Structure: A Mathematically Provable Two-Stage Training Dynamics in Transformers ​

Author: Zixuan Gong, Shijia Li, Yong Liu, Jiaye Teng
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2502.20681v3 Announce Type: replace-cross Abstract: Transformers may exhibit two-stage training dynamics during the real-world training process. For instance, when training GPT-2 on the Counterfact dataset, the answers progress from syntactically incorrect to syntactically correct to semantica...

📖 Read original article


416. SpurLens: Automatic Detection of Spurious Cues in Multimodal LLMs ​

Author: Parsa Hosseini, Sumit Nawathe, Mazda Moayeri, Sriram Balasubramanian, Soheil Feizi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2503.08884v3 Announce Type: replace-cross Abstract: Unimodal vision models are known to rely on spurious correlations, but it remains unclear to what extent Multimodal Large Language Models (MLLMs) exhibit similar biases despite language supervision. In this paper, we investigate spurious bias...

📖 Read original article


417. Measuring AI Ability to Complete Long Software Tasks ​

Author: Thomas Kwa, Ben West, Joel Becker, Amy Deng, Katharyn Garcia, Max Hasin, Sami Jawhar, Megan Kinniment, Nate Rush, Sydney Von Arx, Ryan Bloom, Thomas Broadley, Haoxing Du, Brian Goodrich, Nikola Jurkovic, Luke Harold Miles, Seraphina Nix, Tao Lin, Chris Painter, Neev Parikh, David Rein, Lucas Jun Koba Sato, Hjalmar Wijk, Daniel M. Ziegler, Elizabeth Barnes, Lawrence Chan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2503.14499v4 Announce Type: replace-cross Abstract: Despite rapid progress on AI benchmarks, the real-world meaning of benchmark performance remains unclear. To quantify the capabilities of AI systems in terms of human capabilities, we propose a new metric: 50%-task-completion time horizon. Th...

📖 Read original article


418. Hyperflux: Pruning Reveals Importance ​

Author: Eugen Barbulescu, Antonio Alexoaie, Lucian Busoniu
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2504.05349v5 Announce Type: replace-cross Abstract: Network pruning is used to reduce inference latency and power consumption in large neural networks. However, most methods focus on empirical results at the expense of understanding the pruning process. We introduce Hyperflux, a novel $L_0$ me...

📖 Read original article


419. A computational model of infant sensorimotor exploration in the mobile paradigm ​

Author: Josua Spisak, Sergiu Tcaci Popescu, Stefan Wermter, Matej Hoffmann, J. Kevin O'Regan
Published: 7/14/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG

arXiv:2504.17939v2 Announce Type: replace-cross Abstract: We present a computational model of the mechanisms that may determine infant behavior in the "mobile paradigm". This paradigm has been used in developmental psychology to explore how infants learn the sensory effects of their actions. In this...

📖 Read original article


420. A Provably Convergent Plug-and-Play Framework for Stochastic Bilevel Optimization ​

Author: Tianshu Chu, Dachuan Xu, Wei Yao, Chengming Yu, Jin Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2505.01258v2 Announce Type: replace-cross Abstract: Bilevel optimization has recently attracted significant attention in machine learning due to its wide range of applications and advanced hierarchical optimization capabilities. In this paper, we propose a plug-and-play framework, named PnPBO,...

📖 Read original article


421. Parameter estimation for land-surface models using Neural Physics ​

Author: Ruiyue Huang, Claire E. Heaney, Maarten van Reeuwijk
Published: 7/14/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG

arXiv:2505.02979v4 Announce Type: replace-cross Abstract: We propose a novel inverse-modelling approach that estimates the parameters of a simple land-surface model (LSM) by assimilating data into a differentiable, physics-based forward model formulated using convolutional operations. The governing ...

📖 Read original article


422. Sense and Sensitivity: Examining the Influence of Semantic Recall on Long Context Code Understanding ​

Author: Adam \v{S}torek, Mukur Gupta, Samira Hajizadeh, Prashast Srivastava, Suman Jana
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SE

arXiv:2505.13353v5 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed for understanding large codebases, but whether they understand operational semantics of long code context or rely on pattern matching shortcuts remains unclear. We distinguish between lex...

📖 Read original article


423. Gaussian Invariant Markov Chain Monte Carlo ​

Author: Michalis K. Titsias, Angelos Alexopoulos, Siran Liu, Petros Dellaportas
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2506.21511v2 Announce Type: replace-cross Abstract: We develop sampling methods, which consist of Gaussian invariant versions of random walk Metropolis (RWM), Metropolis adjusted Langevin algorithm (MALA) and second order Hessian or Manifold MALA. Unlike standard RWM and MALA, we show that Gau...

📖 Read original article


424. Generalized and Unified Equivalences between Hardness and Pseudoentropy ​

Author: Lunjia Hu, Salil Vadhan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CC, cs.CR, cs.LG

arXiv:2507.05972v3 Announce Type: replace-cross Abstract: Pseudoentropy characterizations give quantitatively precise formulations of the relationship between computational hardness and computational randomness. We prove a unified pseudoentropy characterization that generalizes and strengthens previ...

📖 Read original article


425. Neural Human Pose Prior ​

Author: Michal Heker, Sefy Kagarlitsky, David Tolpin
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2507.12138v2 Announce Type: replace-cross Abstract: We introduce a principled, data-driven approach for modeling a neural prior over human body poses using normalizing flows. Unlike heuristic or low-expressivity alternatives, our method leverages RealNVP to learn a flexible density over poses ...

📖 Read original article


426. funOCLUST: Clustering Functional Data with Outliers ​

Author: Katharine M. Clark, Paul D. McNicholas
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2508.00110v2 Announce Type: replace-cross Abstract: Functional data present unique challenges for clustering due to their infinite-dimensional nature and potential sensitivity to outliers. An extension of the OCLUST algorithm to the functional setting is proposed to address these issues. The a...

📖 Read original article


427. Detecting and measuring respiratory events in horses during exercise with a microphone: deep learning vs. standard signal processing ​

Author: Jeanne I. M. Parmentier (Utrecht University, University of Twente, Inertia Technology B.V), Rhana M. Aarts (Utrecht University), Elin Hernlund (Swedish University of Agricultural Sciences), Marie Rhodin (Swedish University of Agricultural Sciences), Berend Jan van der Zwaag (University of Twente, Inertia Technology B.V)
Published: 7/14/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2508.02349v2 Announce Type: replace-cross Abstract: Monitoring respiration parameters such as respiratory rate could be beneficial to understand the impact of training on equine health and performance and ultimately improve equine welfare. In this work, we compare deep learning-based methods t...

📖 Read original article


428. Likelihood Matching for Diffusion Models ​

Author: Lei Qian, Wu Su, Yanqi Huang, Song Xi Chen
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.AP, stat.ME, stat.TH

arXiv:2508.03636v3 Announce Type: replace-cross Abstract: We propose a Likelihood Matching approach for training diffusion models by first establishing an equivalence between the likelihood of the target data distribution and a likelihood along the sample path of the reverse diffusion. To efficientl...

📖 Read original article


429. Segmentation and Classification of Pap Smear Images for Cervical Cancer Detection Using Deep Learning ​

Author: Nisreen Albzour, Sarah S. Lam
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2508.17728v3 Announce Type: replace-cross Abstract: Cervical cancer remains a significant global health concern and a leading cause of cancer-related deaths among women. Early detection through Pap smear tests is essential to reduce mortality rates; however, the manual examination is time cons...

📖 Read original article


430. Forecasting Generative Amplification ​

Author: Henning Bahl, Sascha Diefenbacher, Nina Elmer, Tilman Plehn, Jonas Spinner
Published: 7/14/2026, 4:00:00 AM
Categories: hep-ph, cs.LG

arXiv:2509.08048v4 Announce Type: replace-cross Abstract: Generative networks are perfect tools to enhance the speed and precision of LHC simulations. Especially when generating events beyond the size of the training dataset, it is important to understand their statistical precision. We present two ...

📖 Read original article


431. Maximum diversity and weighting for invariants of periodic time series ​

Author: Byungchang So
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, eess.SP, math.MG

arXiv:2509.11146v2 Announce Type: replace-cross Abstract: Magnitude, obtained as a special case of Euler characteristic of enriched category, represents a sense of the size of metric spaces and is related to classical notions such as cardinality, dimension, and volume. While the studies have explain...

📖 Read original article


432. Efficient Group Lasso Regularized Rank Regression with Simulation-Based Tuning ​

Author: Meixia Lin, Mengjiao Shi, Yunhai Xiao, Qian Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC, math.ST, stat.TH

arXiv:2510.11546v3 Announce Type: replace-cross Abstract: High-dimensional regression often suffers from heavy-tailed noise and outliers, which can severely undermine the reliability of least-squares based methods. To improve robustness, we adopt a non-smooth Wilcoxon score based rank objective and ...

📖 Read original article


433. Exact Dynamics of Multi-class Stochastic Gradient Descent ​

Author: Elizabeth Collins-Woodfin, Inbar Seroussi
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC, math.PR

arXiv:2510.14074v2 Announce Type: replace-cross Abstract: We develop a framework for analyzing the learning dynamics of high-dimensional problems trained using one-pass stochastic gradient descent (SGD) with data from multiple anisotropic classes. Our main theorem provides exact expressions for quan...

📖 Read original article


434. Learnable Mixed Nash Equilibria are Collectively Rational ​

Author: Geelon So, Yi-An Ma
Published: 7/14/2026, 4:00:00 AM
Categories: cs.GT, cs.LG

arXiv:2510.14907v2 Announce Type: replace-cross Abstract: We extend the study of learning in games to dynamics that exhibit non-asymptotic stability. We do so through the notion of uniform stability, which is concerned with equilibria of individually utility-seeking dynamics. Perhaps surprisingly, i...

📖 Read original article


435. Survival of the fittest Cox model: Pivotal variable selection for time-to-event data ​

Author: Maxime van Cutsem, Sylvain Sardy
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2510.19374v2 Announce Type: replace-cross Abstract: We revisit Cox's proportional hazards model to improve variable selection in survival analysis. A square-root transformation of the partial likelihood renders the selection of the regularization parameter pivotal, free of the unknown baseline...

📖 Read original article


436. Oracle-Efficient Combinatorial Semi-Bandits ​

Author: Jung-hun Kim, Milan Vojnovi'c, Min-hwan Oh
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2510.21431v2 Announce Type: replace-cross Abstract: We study the combinatorial semi-bandit problem where an agent selects a subset of base arms and receives individual feedback. While this generalizes the classical multi-armed bandit and has broad applicability, its scalability is limited by t...

📖 Read original article


437. Monitoring the calibration of probability forecasts with an application to concept drift detection involving image classification ​

Author: Christopher T. Franck, Anne R. Driscoll, Zoe Szajnfarber, William H. Woodall
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2510.25573v2 Announce Type: replace-cross Abstract: Machine learning approaches for image classification have led to impressive advances in that field. For example, convolutional neural networks are able to achieve remarkable image classification accuracy across a wide range of applications in...

📖 Read original article


438. Redundancy Maximization as a Principle of Associative Memory Learning in Hopfield Networks ​

Author: Mark Bl"umel, Andreas C. Schneider, Valentin Neuhaus, David A. Ehrlich, Marcel Graetz, Michael Wibral, Abdullah Makkeh, Viola Priesemann
Published: 7/14/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, cs.NE, math.IT, physics.comp-ph

arXiv:2511.02584v2 Announce Type: replace-cross Abstract: Associative memory, traditionally modeled by Hopfield networks, enables the retrieval of previously stored patterns from partial or noisy cues. Yet, the local computational principles which are required to enable this function remain incomple...

📖 Read original article


439. Online conformal inference with retrospective adjustment for faster adaptation to distribution shift ​

Author: Jungbin Jun, Ilsang Ohn
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2511.04275v2 Announce Type: replace-cross Abstract: Conformal prediction has emerged as a powerful framework for constructing distribution-free prediction sets with guaranteed coverage assuming only the exchangeability assumption. However, this assumption is often violated in online environmen...

📖 Read original article


440. Towards Blind Lens Aberration Correction via Large LensLib Pre-training and Discrete Degradation Priors ​

Author: Xiaolong Qian, Qi Jiang, Yao Gao, Lei Sun, Kailun Yang, Xian Wang, Zhonghua Yi, Wenyong Li, Ming-Hsuan Yang, Luc Van Gool, Kaiwei Wang
Published: 7/14/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, physics.optics

arXiv:2511.17126v5 Announce Type: replace-cross Abstract: Emerging deep-learning-based lens library pre-training (LensLib-PT) pipeline offers a new avenue for blind lens aberration correction by training a universal neural network, demonstrating strong capability in handling diverse unknown optical ...

📖 Read original article


441. SPQR: A Multi-Dimensional Benchmark for Safety Alignment under Benign Model Adaptation ​

Author: Mohammed Talha Alam, Nada Saadi, Fahad Shamshad, Nils Lukas, Karthik Nandakumar, Fahkri Karray, Samuele Poppi
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CV, cs.LG

arXiv:2511.19558v2 Announce Type: replace-cross Abstract: Text-to-image diffusion models can emit copyrighted, unsafe, or private content. Safety alignment aims to suppress specific concepts, yet evaluations seldom test whether safety persists under benign downstream fine-tuning routinely applied af...

📖 Read original article


442. On the Condition Number Dependency in Bilevel Optimization ​

Author: Lesi Chen, Jingzhao Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: math.OC, cs.AI, cs.LG

arXiv:2511.22331v3 Announce Type: replace-cross Abstract: Bilevel optimization minimizes an objective function, defined by an upper-level problem whose feasible region is the solution of a lower-level problem. We study the oracle complexity of finding an $\epsilon$-stationary point with first-order ...

📖 Read original article


443. Graph-Based Bayesian Optimization for Quantum Circuit Architecture Search with Uncertainty Calibrated Surrogates ​

Author: Prashant Kumar Choudhary, Nouhaila Innan, Muhammad Shafique, Rajeev Singh
Published: 7/14/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG, cs.NE, cs.NI

arXiv:2512.09586v2 Announce Type: replace-cross Abstract: Quantum circuit design is a key bottleneck for practical quantum machine learning on complex, real-world data. We present an automated framework that discovers and refines variational quantum circuits (VQCs) using graph-based Bayesian optimiz...

📖 Read original article


444. An Elementary Proof of the Near Optimality of LogSumExp Smoothing ​

Author: Thabo Samakhoana, Benjamin Grimmer
Published: 7/14/2026, 4:00:00 AM
Categories: math.ST, cs.LG, math.OC, stat.TH

arXiv:2512.10825v3 Announce Type: replace-cross Abstract: We consider the design of smoothings of the (coordinate-wise) max function in $\mathbb{R}^d$ in the infinity norm. The LogSumExp function $f(x)=\ln(\sum^d_i\exp(x_i))$ provides a classical smoothing, differing from the max function in value b...

📖 Read original article


445. NMIRacle: Multi-modal Generative Molecular Elucidation from IR and NMR Spectra ​

Author: Federico Ottomano, Yingzhen Li, Alex M. Ganose
Published: 7/14/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG

arXiv:2512.19733v3 Announce Type: replace-cross Abstract: Molecular structure elucidation from spectroscopic data is a long-standing challenge in Chemistry, traditionally requiring expert interpretation. We introduce NMIRacle, a two-stage generative framework that builds upon recent paradigms in AI-...

📖 Read original article


446. Self-Creating Random Walks for Decentralized Learning under Pac-Man Attacks ​

Author: Xingran Chen, Parimal Parag, Rohit Bhagat, Salim El Rouayheb
Published: 7/14/2026, 4:00:00 AM
Categories: cs.MA, cs.LG

arXiv:2601.07674v2 Announce Type: replace-cross Abstract: Random walk (RW)-based algorithms have long been popular in distributed systems due to low overheads and scalability, with recent growing applications in decentralized learning. However, their reliance on local interactions makes them inheren...

📖 Read original article


447. Accelerated MR Elastography Using Learned Neural Network Representation ​

Author: Xi Peng
Published: 7/14/2026, 4:00:00 AM
Categories: eess.SP, cs.CV, cs.LG, q-bio.QM

arXiv:2601.11878v2 Announce Type: replace-cross Abstract: To develop a deep-learning method for achieving fast high-resolution MR elastography from highly undersampled data without the need of high-quality training dataset. We first framed the deep neural network representation as a nonlinear extens...

📖 Read original article


448. Reinforcement Learning in the Real World: A Survey of Statistical Challenges and Future Directions ​

Author: Asim H. Gazi, Yongyi Guo, Daiqi Gao, Ziping Xu, Kelly W. Zhang, Susan A. Murphy
Published: 7/14/2026, 4:00:00 AM
Categories: stat.AP, cs.LG, stat.ML

arXiv:2601.15353v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has achieved remarkable success in real-world decision-making across diverse domains, including gaming, robotics, online advertising, public health, and natural language processing. Despite these advances, a substa...

📖 Read original article


449. PUMA: Perception-driven Unified Foothold Prior for Mobility Augmented Quadruped Parkour ​

Author: Liang Wang, Kanzhong Yao, Yang Liu, Weikai Qin, Jun Wu, Zhe Sun, Qiuguo Zhu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2601.15995v2 Announce Type: replace-cross Abstract: Parkour tasks for quadrupeds have emerged as a promising benchmark for agile locomotion. While human athletes can effectively perceive environmental characteristics to select appropriate footholds for obstacle traversal, endowing legged robot...

📖 Read original article


450. FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale ​

Author: Ajay Patel, Colin Raffel, Chris Callison-Burch
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2601.22146v2 Announce Type: replace-cross Abstract: Due to limited supervised training data, large language models (LLMs) are typically pre-trained via a self-supervised "predict the next word" objective on a vast amount of unstructured text data. To make the resulting model useful to users, i...

📖 Read original article


451. Bures-Wasserstein Importance-Weighted Evidence Lower Bound: Exposition and Applications ​

Author: Peiwen Jiang, Takuo Matsubara, Minh-Ngoc Tran
Published: 7/14/2026, 4:00:00 AM
Categories: stat.CO, cs.LG, stat.ME

arXiv:2602.04272v2 Announce Type: replace-cross Abstract: The Importance-Weighted Evidence Lower Bound (IW-ELBO) has emerged as an effective objective for variational inference (VI), tightening the standard ELBO and mitigating the mode-seeking behaviour. However, optimizing the IW-ELBO in Euclidean ...

📖 Read original article


452. Reliable Mislabel Detection for Video Capsule Endoscopy Data ​

Author: Julia Werner, Julius Oexle, Oliver Bause, Maxime Le Floch, Franz Brinkmann, Hannah Tolle, Jochen Hampe, Oliver Bringmann
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2602.06938v3 Announce Type: replace-cross Abstract: The classification performance of deep neural networks relies strongly on access to large, accurately annotated datasets. In medical imaging, however, obtaining such datasets is particularly challenging since annotations must be provided by s...

📖 Read original article


453. Precedence-Constrained Decision Trees and Coverings ​

Author: Micha{\l} Szyfelbein, Dariusz Dereniowski
Published: 7/14/2026, 4:00:00 AM
Categories: cs.DS, cs.LG

arXiv:2602.21312v4 Announce Type: replace-cross Abstract: This work considers a number of optimization problems and reductive relations between them. The two main problems we are interested in are the Optimal Decision Tree and Set Cover. We study these two fundamental tasks under precedence constrai...

📖 Read original article


454. MetaState: Persistent Working Memory Enhances Reasoning in Discrete Diffusion Language Models ​

Author: Kejing Xia, Mingzhe Li, Lixuan Wei, Zhenbang Du, Xiangchi Yuan, Dachuan Shi, Qirui Jin, Wenke Lee
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2603.01331v3 Announce Type: replace-cross Abstract: Discrete diffusion language models (dLLMs) generate text by iteratively denoising a masked sequence. However, standard dLLMs condition each denoising step solely on the current hard-masked sequence, while intermediate continuous representatio...

📖 Read original article


455. Context-Dependent Affordance Computation in Vision-Language Models ​

Author: Murad Farzulla
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2603.04419v2 Announce Type: replace-cross Abstract: We characterize the phenomenon of context-dependent affordance computation in vision-language models (VLMs). Our primary study uses Qwen3-VL-30B-A3B ($n = 3{,}213$ scene-context pairs from COCO-2017: 479 images under 7 agentic personas), with...

📖 Read original article


456. Bilateral Trade Under Heavy-Tailed Valuations: Minimax Regret with Infinite Variance ​

Author: Hangyi Zhao
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.GT, cs.LG

arXiv:2603.06851v2 Announce Type: replace-cross Abstract: We study contextual bilateral trade under full feedback when, conditionally on the context, trader valuations have bounded density but infinite variance. We first extend the self-bounding property of Bachoc et al. (ICML 2025) from bounded to ...

📖 Read original article


457. An Approximate Graph Elicits Detonation Lattice ​

Author: Vansh Sharma, Venkat Raman
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, physics.comp-ph, physics.data-an

arXiv:2603.16524v2 Announce Type: replace-cross Abstract: This study presents a novel algorithm based on graph theory for the precise segmentation and measurement of detonation cells from 3D pressure traces, termed detonation lattices, addressing the limitations of manual and primitive 2D edge detec...

📖 Read original article


458. Tokenization vs. Augmentation: A Systematic Study of Writer Variance in IMU-Based Online Handwriting Recognition ​

Author: Jindong Li, Dario Zanca, Vincent Christlein, Tim Hamann, Jens Barth, Peter K"ampf, Bj"orn Eskofier
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG, eess.SP

arXiv:2603.16883v2 Announce Type: replace-cross Abstract: Inertial measurement unit-based online handwriting recognition enables the recognition of input signals collected across different writing surfaces but remains challenged by uneven character distributions and inter-writer variability. In this...

📖 Read original article


459. An Empirical Recipe for Universal Phone Recognition ​

Author: Shikhar Bharadwaj, Chin-Jou Li, Kwanghee Choi, Eunjung Yeo, William Chen, Shinji Watanabe, David R. Mortensen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SD, eess.AS

arXiv:2603.29042v2 Announce Type: replace-cross Abstract: Phone recognition (PR) is a key enabler of multilingual and low-resource speech processing tasks, yet robust performance remains elusive. Highly performant English-focused models do not generalize across languages, while multilingual models u...

📖 Read original article


460. StanceMoE: Mixture-of-Experts Architecture for Stance Detection ​

Author: Abdullah Al Shafi, Md. Milon Islam, Sk. Imran Hossain, K. M. Azharul Hasan
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.00878v2 Announce Type: replace-cross Abstract: Actor-level stance detection aims to determine an author expressed position toward specific geopolitical actors mentioned or implicated in a text. Although transformer-based models have achieved relatively good performance in stance classific...

📖 Read original article


461. Rare Event Analysis via Stochastic Optimal Control ​

Author: Yuanqi Du, Jiajun He, Dinghuai Zhang, Eric Vanden-Eijnden, Carles Domingo-Enrich
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC, physics.chem-ph

arXiv:2604.13213v3 Announce Type: replace-cross Abstract: Rare events such as conformational changes in biomolecules, phase transitions, and chemical reactions are central to the behavior of many physical systems, yet they are extremely difficult to study computationally because unbiased simulations...

📖 Read original article


462. SegWithU: Uncertainty as Perturbation Energy for Single-Forward-Pass Risk-Aware Medical Image Segmentation ​

Author: Tianhao Fu, Austin Wang, Charles Chen, Roby Aldave-Garza, Yucheng Chen
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2604.15271v3 Announce Type: replace-cross Abstract: Reliable uncertainty estimation is critical for medical image segmentation, where automated contours feed downstream quantification and clinical decision support. Many strong uncertainty methods require repeated inference, while efficient sin...

📖 Read original article


463. BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation ​

Author: Baoyou Chen, Hanchen Xia, Peng Tu, Haojun Shi, Liwei Zhang, Yuxuan Yao, Weihao Yuan, Siyu Zhu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2604.16514v5 Announce Type: replace-cross Abstract: Autoregressive vision-language models (VLMs) deliver strong multimodal capability, but their token-by-token decoding imposes a fundamental inference bottleneck. Diffusion VLMs offer a more parallel decoding paradigm, yet directly converting a...

📖 Read original article


464. Fairness Constraints in High-Dimensional Generalized Linear Models ​

Author: Yixiao Lin, James Booth
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2604.16610v2 Announce Type: replace-cross Abstract: Machine learning models often inherit biases from historical data, raising critical concerns about fairness and accountability. Conventional fairness interventions typically require access to sensitive attributes like gender or race, but priv...

📖 Read original article


465. Algorithm Selection with Zero Domain Knowledge via Text Embeddings ​

Author: Stefan Szeider
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2604.19753v2 Announce Type: replace-cross Abstract: We propose a feature-free approach to algorithm selection: instead of hand-crafted instance features, we use pretrained text embeddings. Our method, ZeroFolio, proceeds in three steps. First, it reads the raw instance file as plain text. Seco...

📖 Read original article


466. Ideological Bias in LLMs' Economic Causal Reasoning ​

Author: Donggyu Lee, Hyeok Yun, Jungwon Kim, Junsik Min, Sungwon Park, Sangyoon Park, Jihee Kim
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CE, cs.CL, cs.LG, econ.GN, q-fin.EC

arXiv:2604.21334v2 Announce Type: replace-cross Abstract: Do large language models (LLMs) exhibit systematic ideological bias when reasoning about economic causal effects? As LLMs are increasingly used in policy analysis and economic reporting, where directionally correct causal judgments are essent...

📖 Read original article


467. Recursive Multi-Agent Systems ​

Author: Jiaru Zou, Rui Pan, Ruizhong Qiu, Pan Lu, Shizhe Diao, Jindong Jiang, Hanghang Tong, Tong Zhang, Markus J. Buehler, Jingrui He, James Zou
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2604.25917v2 Announce Type: replace-cross Abstract: Recursive or looped language models have recently emerged as a new scaling axis by iteratively refining the same model computation over latent states to deepen reasoning. We extend such scaling principle from a single model to multi-agent sys...

📖 Read original article


468. Community-Aware Vertex Ordering for Reference-Based Graph Compression: A Cross-Encoder Empirical Study ​

Author: Jimmy Dubuisson
Published: 7/14/2026, 4:00:00 AM
Categories: cs.SI, cs.LG

arXiv:2605.21510v2 Announce Type: replace-cross Abstract: Reference-based graph compression encodes each vertex's neighbor list as differences from a nearby encoded list. WebGraph's BVGraph fixes a single encoding pipeline and relies on a separately chosen vertex ordering -- typically URL-lexicograp...

📖 Read original article


469. Branched Signature Kernel Solvers for ODEs with rough Single-Trajectory signals ​

Author: Munawar Ali, Qi Feng, Charlie Pyle, George Xu
Published: 7/14/2026, 4:00:00 AM
Categories: math.NA, cs.CE, cs.LG, cs.NA

arXiv:2605.25826v2 Announce Type: replace-cross Abstract: We develop a branched signature kernel solver for linear and nonlinear ordinary differential equations driven by a \emph{single observed trajectory} of a possibly rough forcing signal--a setting common within earthquake engineering, finance, ...

📖 Read original article


470. Let It Be Simple: One-Step Action Generation for Vision-Language-Action Models ​

Author: Yitong Chen, Shiduo Zhang, Jingjing Gong, Xipeng Qiu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.RO

arXiv:2606.05737v2 Announce Type: replace-cross Abstract: Generating diverse images from sparse text is hard; generating compact actions from rich observations is easier. From the condition-target view, Vision-Language-Action (VLA) thus aligns with image-to-text, not text-to-image. We formalize this...

📖 Read original article


471. When Does Delegation Beat Majority? A Delegation-Based Aggregator for Multi-Sample LLM Inference ​

Author: Yasushi Sakai, Allen Song, Kent Larson
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.08098v3 Announce Type: replace-cross Abstract: Majority voting is the default unsupervised aggregator for multi-sample LLM inference, but it discards two signals: within-group answer entropy and between-group reasoning geometry. We aggregate by delegation instead (Propagational Proxy Voti...

📖 Read original article


472. Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models ​

Author: Yifu Yuan, Yaoting Huang, Xianze Yao, Yutong Li, Shuoheng Zhang, Linqi Han, Pengyi Li, Jiangeng Sun, Wenting Jia, Zhao Zhang, Yuhao Liu, Ruihao Liao, Yucheng Hu, Qiyu Wu, Yuxiao Li, Zibin Dong, Fei Ni, Yan Zheng, Shuyang Gu, Yi Ma, Hongyao Tang, Han Hu, Jianye Hao
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2606.11324v2 Announce Type: replace-cross Abstract: We introduce Embodied-R1.5, a unified Embodied Foundation Model (EFM) that integrates comprehensive embodied reasoning capabilities, spanning embodied cognition, task planning, correction, and pointing, within a single architecture toward gen...

📖 Read original article


473. ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories ​

Author: Siyuan Luo, Nairong Zheng, Lin Zhou, Tiankuo Yao, Shengyou Yuan, Haojia Yu, Cong Pang, Jiapeng Luo, Lewei Lu
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2606.11520v4 Announce Type: replace-cross Abstract: Training capable OS agents requires data that simultaneously captures structured user intents, multi-turn task delegation, and grounded tool execution--properties absent from existing datasets. We propose ISE (Intent -> Simulate -> Execute), ...

📖 Read original article


474. KVEraser: Learning to Steer KV Cache for Efficient Localized Context Erasing ​

Author: Mufei Li, Shikun Liu, Dongqi Fu, Haoyu Wang, Yinglong Xia, Hong Li, Hong Yan, Pan Li
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2606.17034v2 Announce Type: replace-cross Abstract: Post-hoc context erasing over the KV cache is challenging because a local edit has a global consequence: once a span has been processed, its influence propagates into the cached states of all subsequent tokens. This issue arises naturally in ...

📖 Read original article


475. Betting on Moments: Legendre Jumper Martingales for Online Exchangeability Testing ​

Author: Johan Hallberg Szabadv'ary
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2606.20859v2 Announce Type: replace-cross Abstract: A fundamental assumption in statistics and machine learning is that ``the future looks like the past,'' formalized as exchangeability: the joint data distribution is order-invariant. In practice, this assumption is often violated due to distr...

📖 Read original article


476. Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction ​

Author: Chenguang Wang, Ming Li, Xinyue Zeng, Zhuochun Li, Hong Jiao, Tianyi Zhou, Dawei Zhou
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG

arXiv:2606.28186v2 Announce Type: replace-cross Abstract: Predicting human item difficulty is central to educational assessment, where reliable estimates support fairness and effective test construction. Existing methods often depend on costly human calibration or item-level textual representations,...

📖 Read original article


477. Safety from Honesty in a Disinterested AI Predictor ​

Author: Yoshua Bengio, Oliver Richardson, Tom'a\v{s} Gaven\v{c}iak, Michael Cohen, Rory Svarc, Damiano Fornasiere, Gael Gendron, David Hyland, Aton Kamanda, Adam Oberman, Francis Rhys Ward, Anna Gaven\v{c}iak, Jacob Livingston Slosser, Vincent Mai, Iulian Serban, Joumana Ghosn
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.29657v2 Announce Type: replace-cross Abstract: As AI systems become more capable, training procedures that optimize for downstream outcomes risk introducing implicit agency: goal-directed behavior that designers never specified. We present a formal safety argument for the Scientist AI (SA...

📖 Read original article


478. Freeform Preference Learning for Robotic Manipulation ​

Author: Marcel Torne, Anubha Mahajan, Abhijnya Bhat, Chelsea Finn
Published: 7/14/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2606.32027v2 Announce Type: replace-cross Abstract: Reward design remains a central bottleneck for autonomous robot policy improvement, especially in long-horizon manipulation tasks where sparse success labels provide too little signal and binary preferences collapse many competing notions of ...

📖 Read original article


479. Optimal scaling of MCMC algorithms: the Hamiltonian approach ​

Author: P. Dobson, J. M. Sanz-Serna, K. C. Zygalakis
Published: 7/14/2026, 4:00:00 AM
Categories: stat.CO, cs.LG, math.PR

arXiv:2607.00586v2 Announce Type: replace-cross Abstract: We present a simple, yet general approach to study the scaling properties as the dimensionality of Metropolised MCMC sampling algorithms increases. The study relies on the symmetries of the Hamiltonian formalism and ultimately on the symmetry...

📖 Read original article


480. SUNTA: Hierarchical Video Prediction with Surprise-based Chunking ​

Author: Tomoshi Iiyama, Masahiro Suzuki, Yutaka Matsuo
Published: 7/14/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.02087v2 Announce Type: replace-cross Abstract: Hierarchical state-space models (HSSMs) offer a promising approach to long-horizon prediction by segmenting sequences into temporal chunks. However, their performance hinges on how chunk boundaries are determined. While prior HSSMs typically ...

📖 Read original article


481. LRX-PINN: A Layer-Resolving XNet Physics-Informed Neural Network with Integrated Cauchy Activations for Convection-Dominated Problems ​

Author: Zihao Guo, Xin Li, Zhihong Xia
Published: 7/14/2026, 4:00:00 AM
Categories: math.AP, cs.LG

arXiv:2607.03682v2 Announce Type: replace-cross Abstract: Convection-dominated convection-diffusion problems often develop thin layers, where the solution has sharp transition profiles and its derivatives are highly localized. This creates a structural mismatch for standard physics-informed neural n...

📖 Read original article


482. Geometric Causal Models ​

Author: Eli N. Weinstein, David M. Blei
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, q-bio.BM

arXiv:2607.05153v2 Announce Type: replace-cross Abstract: Scientists often seek to draw causal inferences from structured data that is not independently and identically distributed, such as spatial data, network data, or molecular data. We develop geometric causal models (GCMs), a framework for caus...

📖 Read original article


483. Is the Geometry Doing the Work? An Operating-Point Audit of Hierarchy in Hyperbolic Vision-Language Models ​

Author: Jaeyoung Kim, Eunseok Kim, Dongsuk Jang
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.05268v2 Announce Type: replace-cross Abstract: Whether a hyperbolic representation model uses its geometry cannot be inferred from curvature alone: what matters is the dimensionless operating point $\sqrt{c}\rho$ and whether the radial and cone mechanisms are operational there. We develop...

📖 Read original article


484. Optimization Geometrodynamics: Variational Reduction and Interaction Curvature ​

Author: Zavier Li
Published: 7/14/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DG

arXiv:2607.06723v2 Announce Type: replace-cross Abstract: Adaptive optimizers carry hidden states that change how visible gradients become parameter motion. We develop optimization geometrodynamics as a variational theory of this hidden geometry. Infimal pushforward eliminates all hidden states real...

📖 Read original article


485. Restricted Dynamic Geometric Complexity: Path-Space Reduction and M\"obius--Jacobi Response ​

Author: Zavier Li
Published: 7/14/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2607.07204v2 Announce Type: replace-cross Abstract: Structured preconditioners restrict optimization to a small family of positive metrics, but endpoint condition-number reachability does not measure the geometric effort required to reach a useful metric. We formulate this effort as a path-spa...

📖 Read original article


486. Statistical Efficiency and Inference of Quantile Distributional Reinforcement Learning ​

Author: Zijie Cheng, Yang Peng, Zhihua Zhang
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.08444v2 Announce Type: replace-cross Abstract: In this paper, we study quantile-based distributional reinforcement learning from the perspective of statistical efficiency. We focus on distributional policy evaluation, whose goal is to characterize the return distribution, namely the distr...

📖 Read original article


487. EHR-MPC: Inference-Time Control for Sepsis Treatment with Generative Patient Digital Twins ​

Author: Joshua Pickard, Wei Qi, Na Li, Ann Woolley, Lisa Cosimi, Roy Kishony, Deborah Hung
Published: 7/14/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, cs.SY, eess.SY, math.OC

arXiv:2607.08793v2 Announce Type: replace-cross Abstract: Sepsis is a leading cause of mortality, yet optimal treatment policies remain contested. Existing reinforcement learning (RL) approaches learn fixed strategies for sepsis treatment, limiting adaptability to changing clinical objectives during...

📖 Read original article


488. A Sovereign, Open-Source Foundation Model for German and English ​

Author: The Soofi-Team, :, Benedikt Droste, David Fitzek, Ruben H"arle, Lukas Helff, Maximilian Idahl, Alex Jude, Abbas Goher Khan, Maurice Kraus, Timm Ruland, Richard Rutmann, Sebastian Sztwiertnia, Markus Frey, Daniil Gurgurov, Jan Pfister, Tom R"ohr, Sebastian von Rohrscheidt, J"org Bienert, Nicolas Flores-Herr, Simon Gottschalk, Andreas Hotho, Kristian Kersting, Joachim K"ohler, Alexander L"oser, Wolfgang Nejdl, Simon Ostermann, Jan Plogsties, Patrick Putzky, Mehdi Ali, Michael Fromm, Max L"ubbering
Published: 7/14/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.09424v2 Announce Type: replace-cross Abstract: We present Soofi S 30B-A3B, a sovereign, open-source Mixture-of-Experts (MoE) hybrid Mamba Transformer foundation model for German and English. Its hybrid design activates only 3B of 30B parameters per token and keeps the inference cache near...

📖 Read original article