Skip to content

arXiv cs.LG - 2026-08-06 ​

291 items collected.


1. C$^2$MOE: Consistency and Complementarity-guided Mixture of Experts for Incomplete Multimodal Emotion Learning ​

Author: Yuntao Shou, Tao Meng, Wei Ai, Keqin Li
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04013v1 Announce Type: new Abstract: Recent advances in Multimodal Emotion Recognition in Conversations (MERC) highlight its reliance on complete multimodal inputs. However, real-world data often suffer from missing modalities due to transmission errors or user behavior, severely degradin...

📖 Read original article


2. On Hamming-Lipschitz Type Stability of the Subdominant (Minmax) Ultrametric: Theory and Simple Proofs ​

Author: Alokendu Mazumder, Arnab Roy, Punit Rathore
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04014v1 Announce Type: new Abstract: The subdominant (minmax) ultrametric is a canonical tree-structured summary of a dissimilarity matrix, arising equivalently as the ultrametric induced by single-linkage clustering. While its classical stability theory is usually formulated in $\ell_\in...

📖 Read original article


3. A Trust-region Framework for Moment Estimation ​

Author: Oluwasegun A. Somefun
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SP, eess.SY

arXiv:2608.04026v1 Announce Type: new Abstract: In this paper, we develop a trust-region framework for understanding the behavior of adaptive moment estimation mechanisms, such as \textsc{Adam}, in stochastic gradient optimization. Specifically, in this framework, the magnitude of the update step fo...

📖 Read original article


4. Learning to Resolve Neutron Resonances with Fully Convolutional Neural Networks ​

Author: Nataly R. Panczyk, Athanasios Stamatopoulos, Josef Svoboda, Majdi I. Radaideh
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, physics.app-ph

arXiv:2608.04027v1 Announce Type: new Abstract: This work investigates the feasibility of augmenting traditional R-Matrix codes with a robust machine learning framework for automatically detecting neutron resonances in transmission spectra. Neutron transmission data are often complex and noisy, maki...

📖 Read original article


5. Lindblad-Inspired Multi-Timescale Reservoir Computing with Separable Rotation and Dissipation ​

Author: Jyotiranjan Beuria, Amit Shukla
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET, cs.NE

arXiv:2608.04028v1 Announce Type: new Abstract: Echo-state networks enable efficient temporal learning by fixing the recurrent dynamics and training only a linear readout. However, conventional reservoirs typically accommodate signal mixing, memory retention, and stability within a single random rec...

📖 Read original article


6. An Explainable LLM Agent Layer for Open-World Anomaly Detection in Oil Wells ​

Author: Lucas Gouveia Omena Lopes, Thales Miranda de Almeida Vieira, Eduardo Toledo de Lima Junior, William Wagner Matos Lira
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04041v1 Announce Type: new Abstract: Open-World Learning (OWL) pipelines for oil well anomaly detection have recently been shown to combine autoencoder-based detection, multiclass classification, and Mahalanobis-based novelty detection on the public 3W dataset. These pipelines answer \tex...

📖 Read original article


7. Tactus: Open-Vocabulary Object Recognition from Low-Cost Pressure Arrays ​

Author: Abdul Basit Tonmoy
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.04043v1 Announce Type: new Abstract: Resistive pressure arrays are the cheapest and most widely shipped tactile sensors, yet tactile representation learning has concentrated on optical sensors that image a deforming gel. We present Tactus, an open model that answers text queries from pres...

📖 Read original article


8. Robust and Personalized Federated Learning for Aircraft-Engine Prognostics under Benign and Adversarial Client Heterogeneity ​

Author: Chinmoy Mitra, Md. Mehedi Hasan Nipu, Mohammad Sakib Mahmood, Md. Rakibul Islam, M. F. Mridha
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2608.04045v1 Announce Type: new Abstract: Federated learning (FL) enables aircraft fleet operators to jointly train remaining-useful-life (RUL) models from engine sensor telemetry without sharing raw data. This study examines two complementary challenges: benign heterogeneity, where honest ope...

📖 Read original article


9. Recurrent Residual Quantization: A Progressive Multi-Precision Representation for LLMs ​

Author: Yu Luo, Bo Dong, Wenhua Cheng, Haihao Shen
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04048v1 Announce Type: new Abstract: Serving large language models (LLMs) under diverse deployment constraints requires flexible trade-offs between accuracy, memory footprint, and throughput. However, conventional quantization methods typically require a separate checkpoint for each targe...

📖 Read original article


10. CAMP: A Cycle-Aware Multi-Scale Patch Mixer for Time Series Forecasting ​

Author: Jung Min Choi, Vijaya Krishna yalavarthi, Lars Schmidt-Thieme
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04051v1 Announce Type: new Abstract: Real-world time series are often governed by recurring patterns, but their dominant periods may vary across datasets, forecasting settings, and individual input windows. Existing cycle-aware forecasters commonly rely on a single period selected at the ...

📖 Read original article


11. LaPrune: Controllable Differentiable Sparsity at Million Scale ​

Author: Jakub Antczak, Joanna Wojciechowicz, {\L}ukasz Struski, Jacek Tabor
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04057v1 Announce Type: new Abstract: Top-$k$ selection determines which components of a sparse model remain active. Hard selection blocks gradients, while continuous relaxations often couple mask hardness to the selected mass. We introduce LaPrune, a mathematically exact-budget differenti...

📖 Read original article


12. SJEPA: Learning Elegant Latent Dynamics with Hybrid Symbolic-Neural Predictors ​

Author: Yongchao Huang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04060v1 Announce Type: new Abstract: Joint-embedding predictive architectures learn abstract states by predicting target embeddings from context embeddings, but their transition models are typically opaque neural maps. We introduce SJEPA, a reconstruction-free JEPA framework that learns p...

📖 Read original article


13. Spend Bits Where Queries Look: KV Cache Vector Quantization with Attention-Preserving Transforms ​

Author: Samuel Fern'andez-Mendui~na, Amir Ziashahabi, Eduardo Pavez, Antonio Ortega, Salman Avestimehr
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, eess.SP, math.IT

arXiv:2608.04074v1 Announce Type: new Abstract: Long-context LLM decoding reads the key-value (KV) cache at every step. Loading it takes longer than computing attention over it, so throughput is bandwidth-bound. Hence, reducing the cache size can raise both decoding speed and serving capacity. The c...

📖 Read original article


14. Spatiotemporal Graph Transformer for Traffic Intelligence in Edge Computing ​

Author: Laha Ale, Letian Lin, Na Cao, Zheng Ma, Peng Yu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04075v1 Announce Type: new Abstract: Accurate traffic forecasting is essential for proactive resource management in edge computing, where service demand evolves dynamically across both space and time. In practical cellular edge systems, traffic exhibits strong spatial correlations among n...

📖 Read original article


15. SpecDrop: Parameter-Free Category-Conditioned Routing for Modular Specialization ​

Author: Boyao Wang, Zhihan Lei
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CV

arXiv:2608.04084v1 Announce Type: new Abstract: Mixture-of-experts (MoE) networks pursue specialization through learned routers, gates, and load-balancing losses, yet at matched total-parameter budgets learned routers can underperform equal-weight No-Routing baselines. Is the bottleneck the routing ...

📖 Read original article


16. Out-Of-The-Loop Multi-Fidelity Bayesian Optimization ​

Author: Gustavo Sutter, Hao Wang, Luis Ricardez-Sandoval, Pascal Poupart, Agustinus Kristiadi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04113v1 Announce Type: new Abstract: Black-box optimization is a ubiquitous problem in science and engineering, often dealing with expensive objective functions with cheaper lower-fidelity proxies available. Multi-fidelity Bayesian optimization (MF-BO) is a principled approach to this pro...

📖 Read original article


17. LiNC: Lightweight Noise Correction via Per-Sample Trust and Gaussian Mixture Modeling ​

Author: Abhishek Moturu, Babak Taati, Anna Goldenberg
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.04147v1 Announce Type: new Abstract: Label noise is common in medical imaging datasets due to factors such as inter-rater variability, annotation errors, and ambiguous cases. This can severely undermine the reliability and clinical effectiveness of machine learning models trained using th...

📖 Read original article


18. MINT: Tensor Decomposition on Stacked Recurrence Matrices for Time Series Data Mining ​

Author: Kaamil Kaka, Audrey Der, Evangelos E. Papalexakis, Zachary Zimmerman, Vikram Jayaram
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04157v1 Announce Type: new Abstract: Recurrence plots are a time series data mining primitive applied to a variety of domains (e.g. star light curves, sound waveforms, CCT telemetry). This work proposes tensorized self-similarity matrices as a primitive for univariate time series datasets...

📖 Read original article


19. Understanding Fault Tolerance of Adversarially Robust Pruned Models ​

Author: Manali Dangarikar, Cory Merkel
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.04173v1 Announce Type: new Abstract: Deep neural networks (DNNs) deployed on resource-constrained neuromorphic hardware face three concurrent challenges: the need for model compression through pruning, vulnerability to adversarial input perturbations, and susceptibility to hardware-induce...

📖 Read original article


20. TS2TabPFN: Time Series Classification and Extrinsic Regression through Feature Extraction and a Tabular Foundation Model ​

Author: Gabriel da Costa Merlin, Diego Furtado Silva
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04174v1 Announce Type: new Abstract: Time series data are ubiquitous in practical applications, where classification (TSC) and extrinsic regression (TSER) have emerged as essential tasks for obtaining value from temporal sequences. While the literature has seen significant progress throug...

📖 Read original article


21. A Comparative Study of Feature Selection Methods for EHR Diagnosis Codes in Opioid Use Disorder Prediction ​

Author: Zihan Ding, Yinan Liu, Tengfei Ma, Rachel Wong, George Leibowitz, Benjamin Littenberg, Xia Zheng, Richard N. Rosenthal, Fusheng Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04180v1 Announce Type: new Abstract: Feature selection is a critical step in electronic health record (EHR)-based predictive modeling, where input variables are often high-dimensional, sparse, noisy, and redundant. Large feature sets not only increase computational burden and overfitting ...

📖 Read original article


22. Unscented KalmanNet: a hybrid deep learning filter with calibrated posterior covariance for nonlinear state estimation ​

Author: Minhyeok Ko, Abdollah Shafieezadeh
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, eess.SP, math.NA

arXiv:2608.04201v1 Announce Type: new Abstract: State estimation for nonlinear dynamical systems is commonly performed with the Unscented Kalman filter (UKF), which propagates the state moments through deterministic sigma points and reports a posterior covariance at every step. In practice, however,...

📖 Read original article


23. From Non-Convex Self-Concordant Regularization to Scalable Quasi-Newton Training of PINNs ​

Author: Chenhao Si, Kang An, Shiqian Ma, Ming Yan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04206v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) often require high-accuracy quasi-Newton refinement to obtain reliable partial differential equation solutions, but their residual objectives can exhibit indefinite, nearly singular, and poorly scaled local curv...

📖 Read original article


24. Attention-Only White-Box Transformer via LeJEPA-Based Self-Supervised Pretraining ​

Author: Yang Bai, Linyuan Wang, Haoyang Jiang, Nuolin Sun, Libin Hou, Bin Yan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04213v1 Announce Type: new Abstract: Existing studies on self-supervised learning for white-box networks typically decouple the derivation of white-box networks via optimization algorithms from self-supervised learning paradigms. In this work, we instead revisit the two components from a ...

📖 Read original article


25. Random features for Grassmannian kernel approximation with bounded rank-one projections ​

Author: R'emi Delogne, Laurent Jacques
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2608.04227v1 Announce Type: new Abstract: We propose a family of random feature maps for scalable kernel machines on low-dimensional subspaces, ie on the Grassmannian manifold. Such representations are useful when data classes or clusters are well described by the span of a few samples. Classi...

📖 Read original article


26. Transferable Dual-Stream Representations for Mesoscale-Preserving Sea Surface Temperature Downscaling ​

Author: Parth Doshi, Priyanka Aravindan, Vaishnav Vaidheeswaran, Md Mahbub Alam, Gabriel Spadon
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.04230v1 Announce Type: new Abstract: Deep learning models for scientific spatio-temporal downscaling often minimize reconstruction error while failing to preserve physically meaningful multi-scale structure. For sea surface temperature prediction, this can yield outputs that are numerical...

📖 Read original article


27. Physics-informed reduced-order modelling with equivariant spectral submanifolds ​

Author: Georg Maierhofer
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.DS, math.NA

arXiv:2608.04239v1 Announce Type: new Abstract: Spectral submanifold (SSM) reduction has emerged as a mathematically principled route to reliable nonlinear reduced-order models, capturing dynamics beyond the reach of linear techniques such as Dynamic Mode Decomposition (DMD). The computation of SSMs...

📖 Read original article


28. Attention-based representations for multi-task computation ​

Author: Daniel Hsu, Mingyue Xu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04243v1 Announce Type: new Abstract: Multi-head attention layers produce vector representations that support multiple downstream tasks. We establish bounds on the number of heads required in two simple and concrete multi-task scenarios. In the first scenario, a vector representation is so...

📖 Read original article


29. Geometry-Informed Parameter-Efficient Fine-Tuning of Pre-trained Molecular GNNs for Blood-Brain Barrier Permeability Prediction ​

Author: Marco Vieto Vega, Long D. Nguyen, Binh P. Nguyen
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04257v1 Announce Type: new Abstract: Blood-brain barrier permeability (BBBP) prediction is a critical screening task in central nervous system drug discovery, where candidate molecules must be assessed for whether they can cross, or should be prevented from crossing, the blood-brain barri...

📖 Read original article


30. Sample Complexity of Multicalibration for Multilevel Properties ​

Author: Jiuyao Lu, Krishnakumar Balasubramanian, Aleksandr Podkopaev, Shiva Prasad Kasiviswanathan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH

arXiv:2608.04288v1 Announce Type: new Abstract: Calibration requires a predictor to be unbiased after conditioning on its own predictions. Multicalibration asks for this guarantee simultaneously across a collection of groups. Many prediction tasks ask for several related features of the same conditi...

📖 Read original article


31. Adaptive Finite-Budget Training for CVaR Risk-Aware Q-Learning ​

Author: Yifan Wu, Junjie Lei, Wenjie Huang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, q-fin.RM

arXiv:2608.04305v1 Announce Type: new Abstract: Risk-aware Q-learning (RaQL) provides a model-free, two-timescale estimator for dynamic risk objectives, but its finite-budget behavior remains fragile: fixed inner-loop hyperparameters can produce unstable value estimates, persistent Bellman residuals...

📖 Read original article


32. ArborEnum: Decision Tree Rashomon Sets over Continuous Features ​

Author: Zakk Heile, Hayden McTavish, Margo Seltzer, Cynthia Rudin
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.04310v1 Announce Type: new Abstract: The Rashomon effect describes the phenomenon that many models can achieve nearly equivalent performance on the same learning task, with significant ramifications for robustness, feature importance, and customizability. These use cases motivate the comp...

📖 Read original article


33. Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits ​

Author: Bo Xue, Ji Cheng, Haodong Jing, Hongzong Li, Shuang Qiu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04324v1 Announce Type: new Abstract: This paper studies generalized low-rank matrix bandits with multiple prioritized objectives. At each round, the learner selects a matrix-valued arm and observes a vector-valued reward, whose components correspond to multiple objectives with different p...

📖 Read original article


34. Real-time probabilistic tsunami forecasting via generative AI ​

Author: Yusuke Oishi, Takashi Furumura, Fumihiko Imamura
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04327v1 Announce Type: new Abstract: Explicit onshore tsunami inundation forecasting can improve public risk awareness, but deterministically predicted inundation boundaries under highly uncertain conditions, such as near-field tsunamis generated by megathrust earthquakes, may falsely imp...

📖 Read original article


35. Cost-Aware Multi-Objective Bandits: Theory and Application to Budgeted LLM Configuration Evaluation ​

Author: Bo Xue, Zhi Hong, Jiayi Li, Yuanyu Wan, Ji Cheng, Shuang Qiu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04333v1 Announce Type: new Abstract: Large language model (LLM) configuration evaluation is challenging due to limited evaluation budgets, varying costs, and multiple competing objectives. In this paper, we formulate LLM configuration evaluation as a cost-aware multi-objective bandit prob...

📖 Read original article


36. ATLAS: Adaptive Topological Learning with Abstract Successors for Continual Learning ​

Author: R. Blake Lawlor, Daniel S. Brown
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04334v1 Announce Type: new Abstract: Contemporary model-free reinforcement learning algorithms can achieve very high performance, but have low sample efficiency and are not robust to changes in the environment. Model-based algorithms have much higher sample efficiency, but still fail when...

📖 Read original article


37. Looking in the Mirror: Introspecting Side-Effect Misalignments Induced by Fine-Tuning ​

Author: Kotaro Yoshida, Laura Gomezjurado Gonzalez, Yukinori Yamamoto, Yuji Naraki, Ryotaro Shimizu, Wenya Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04347v1 Announce Type: new Abstract: Fine-tuning enables a source model to acquire desired capabilities and behaviors in a target domain while retaining much of its general-purpose competence. However, this adaptation process can also degrade alignment properties that were present in the ...

📖 Read original article


38. Manipulation-Proof Oblivious Audits against Deceptive Model Providers ​

Author: Augustin Godinot, Sofiane Azogagh, Julien Ferry, S'ebastien Gambs
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.CY

arXiv:2608.04365v1 Announce Type: new Abstract: Audits have emerged as a critical instrument for algorithmic governance, providing a mechanism for external scrutiny and governance of machine learning models. However, ensuring the integrity of such assessments remains a challenging issue. For instanc...

📖 Read original article


39. EvtGraph: Event-Adaptive Compression for Sparse Temporal Graph Learning in Multimodal Time Series ​

Author: Ziqian Wang, Tingxiong Xiao, Yuxiao Cheng, Jinli Suo
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04368v1 Announce Type: new Abstract: Multimodal temporal data are inherently irregular and uneven in information density, yet most models rely on uniform discretization, leading to inefficient representations. We propose \textbf{EvtGraph}, a unified framework that aligns computation with ...

📖 Read original article


40. Towards Trustworthy Hypergraph Neural Networks under Label Noise ​

Author: Mengyao Zhou, Zhiheng Zhou, Xiao Han, Guiying Yan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04377v1 Announce Type: new Abstract: Hypergraph neural networks (HGNNs) have demonstrated remarkable capabilities in processing complex higher-order relationships. However, their performance is highly dependent on labeled data, making them vulnerable to label noise. Despite advances in le...

📖 Read original article


41. NodeJEPA: Structure-Conditioned Latent Prediction for Node-Level Graph Self-Supervised Learning ​

Author: Tinghe Zhang, Jian Xu, Jiaheng Chen, Jiaxing Li, Yucheng Xiao, Qiang Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04381v1 Announce Type: new Abstract: Self-supervised learning on graphs is largely shaped by contrastive methods that depend on carefully designed augmentations, and by generative methods that reconstruct node attributes in the input space. Both paradigms can entangle representations with...

📖 Read original article


42. Non-asymptotic implicit bias of logistic regression at early-stage gradient descent dynamics ​

Author: Han Bao
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.04382v1 Announce Type: new Abstract: Gradient descent has been of particular interest in modern machine learning beyond sole focus on optimization. Implicit bias emerging from optimization, though not being encoded by the learning objective, often prevents from overfitting to spurious pat...

📖 Read original article


43. NeuroPB: Scaling Neural Decoding with Pretrained Behavioral Representations ​

Author: Luyao Jin, Yonghao Song, Huan Zhao, Vincent C. K. Cheung, Wei-Hsin Liao
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04389v1 Announce Type: new Abstract: Decoding continuous motor trajectories from neural activity is essential for developing practical brain-computer interfaces (BCIs). However, current neural decoders are constrained by the limited scale and heterogeneity of neural recordings. In contras...

📖 Read original article


44. When Proxy Prediction Becomes Equation Reconstruction: Diagnostics and Residual Learning for Factor-Derived Proxy Supervision ​

Author: Chayan Lahiri, Ahmed Shafee, Cody Fehringer
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04393v1 Announce Type: new Abstract: Scientific machine learning often relies on proxy targets computed from known domain factors when direct observations are limited. When those same factors are used as model inputs, however, high predictive accuracy may reflect reconstruction of the pro...

📖 Read original article


45. Elbow-Based MoE Routing: A Training-Free Inference Time Plugin for Expert Selection ​

Author: Robin Pan, Raymond Liu, Daniel Fang, Adelina Andrei, Rosa Wu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04401v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models enable model scaling while maintaining low inference-time compute by activating only a subset of experts per token. However, conventional routing relies on a fixed top-k selection, forcing the model to spend the same com...

📖 Read original article


46. Training-Free Hashing-Based Attention via Binary Principal Components ​

Author: Daohai Yu, Zhanpeng Zeng, Keyu Chen, Wenhao Li, Zhifeng Shen, Luxi Lin, Ruizhi Qiao, Xing Sun, Rongrong Ji
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.04405v1 Announce Type: new Abstract: Long-context large language models (LLMs) are increasingly deployed in real-world applications, yet self-attention remains a major efficiency bottleneck -- especially during decoding -- due to the necessity of repeatedly processing ever-growing key-val...

📖 Read original article


47. MESH: Memory-Efficient Sinkhorn Optimization for Mixture-of-Experts Training ​

Author: Masato Fujitake
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.04407v1 Announce Type: new Abstract: Memory-efficient matrix optimizers such as Sinkhorn gradient descent remove most AdamW optimizer state for dense Transformer matrices, but direct application to Mixture-of-Experts (MoE) training is unreliable. We study this failure in a controlled 110M...

📖 Read original article


48. Not Every Divergence Should Be Suppressed: Counterfactual Recoverability in On-Policy Distillation ​

Author: De Jiang, Zhengyang Zhang, Kehong Yuan, Shaohua Ma
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04408v1 Announce Type: new Abstract: On-policy distillation (OPD) supervises student-visited trajectories, yet divergence-based rules cannot determine whether an erroneous prefix remains correctable. We formulate this decision as counterfactual recoverability and replay each error state t...

📖 Read original article


49. SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation ​

Author: Zikun Qu, Min Zhang, Mingze Kong, Zhiwei Shang, Yikun Ban, Shuang Qiu, Zhongxiang Dai
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04419v1 Announce Type: new Abstract: On-policy distillation (OPD) provides dense teacher supervision on student-generated trajectories, but standard reverse-KL training can assign insufficient probability to other plausible continuations. Teacher entropy alone does not reveal whether unce...

📖 Read original article


50. Generative Optimization for Incentivized Advertising with Global Level Constraints ​

Author: Gege Chen, Ning Luo, Hao Jiang, Da Li, Wenzheng Shu, Teng Sha, Yanxiang Zeng, Wenxin Tai, Fan Zhou, Xialong Liu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04421v1 Announce Type: new Abstract: Incentivized advertising allocates monetary or virtual rewards to drive user engagement, where a key challenge is optimizing continuous incentive magnitudes under strict global constraints. This problem is complicated by high-frequency interactions, de...

📖 Read original article


51. Robustness Emerges Early in Training Dynamics, but Is Not Preserved ​

Author: Jiangang Yang, Wenhui Shi, Lu Hu, Jing Xing, Jian Liu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.04442v1 Announce Type: new Abstract: Robustness to natural corruptions remains a fundamental challenge for deep neural networks. In this paper, we identify a robustness fading phenomenon where shallow layers spontaneously develop robust representations and flat loss landscapes in early tr...

📖 Read original article


52. Tropical Algebraic Geometry for Neuronal Representations: An Arakelov-Green Measure Based Descriptor for Graph Learning ​

Author: Yuyang Zhang, Weihan Xu, Xuehai Zhou, Shucheng Cao, Qihuang Zhang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CG

arXiv:2608.04460v1 Announce Type: new Abstract: The quantitative analysis of 3D neuronal morphologies requires capturing both graph topology and spatial geometry. Current message-passing Graph Neural Networks (GNNs) are bounded by the 1-Weisfeiler-Lehman (1-WL) test, limiting their ability to captur...

📖 Read original article


53. Beyond Linear Dynamics: Neural Bilinear Dynamical Models for Time Series Forecasting ​

Author: Mengzhou Gao, Huangqian Yu, Pengfei Jiao
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04471v1 Announce Type: new Abstract: Time series in real-world applications are often generated by nonlinear dynamical systems, making accurate forecasting challenging. Existing approaches that explicitly model system dynamics typically rely on linear assumptions or Koopman-based lineariz...

📖 Read original article


54. Local Violation Certification for Linear Predict-Then-Optimize Pipelines ​

Author: \c{S}. .Ilker Birbil, Wenhao Chi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.04474v1 Announce Type: new Abstract: Data-driven decision pipelines combining predictive machine learning models with downstream optimization software are increasingly used to make high-stakes operational decisions. Certifying the safety, fairness, and reliability of these decisions is es...

📖 Read original article


55. Discretization and Statistical Consistency of Functional Flow Matching ​

Author: Lennon J. Shikhman
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, math.PR, stat.ML

arXiv:2608.04531v1 Announce Type: new Abstract: Functional flow matching is posed on distributions of functions but implemented from finitely many coefficients or point values. Under scattered or adaptive refinement, the resulting conditioning sigma-algebras need not be nested, so martingale converg...

📖 Read original article


56. Learning Compression Rules for Network Traffic ​

Author: Quentin Lampin (Orange Research), 'Eloi Sainte-Beuve (Orange Research, Universit'e Grenoble Alpes), Louis-Adrien Dufr`ene (Orange Research), Guillaume Larue (Orange Research), Massih-Reza Amini (Universit'e Grenoble Alpes)
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.NI

arXiv:2608.04545v1 Announce Type: new Abstract: We study the problem of learning compact rule-based compressors for structured network traffic. Each packet is a record of header fields that are highly redundant within a flow, and a compressor is a small set of rules matching such records and replaci...

📖 Read original article


57. A Model Merging Approach for Continual MLLM Unlearning ​

Author: Yuhang Wang, Linlin Zhang, Haoxuan Ji, Xianmin Ye, Zhenxing Niu, Haichang Gao
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04548v1 Announce Type: new Abstract: Multimodal large language model (MLLM) unlearning methods have been proposed to remove private, sensitive, or proprietary information from well-trained models. However, most existing MLLM unlearning methods are designed for one-shot requests and fail t...

📖 Read original article


58. Rethinking Reservoir Pruning: A Dynamical Perspective for Echo State Networks ​

Author: Sudip Laudari, Puspa Raj Adhikari
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DS

arXiv:2608.04593v1 Announce Type: new Abstract: Echo State Networks (ESNs) offer an efficient framework for temporal prediction, but their randomly initialized reservoirs are often over-parameterized and dynamically redundant. Existing pruning methods largely rely on static connectivity or activatio...

📖 Read original article


59. Why Ranking Anomaly Detection Algorithms Isn't as Reliable as You May Think ​

Author: Simon Kl"uttermann, J'er^ome Rutinowski, Frederik Polachowski, Alice Kirchheim
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04613v1 Announce Type: new Abstract: Anomaly detection is a safety-critical machine learning problem with applications ranging from fraud detection to network intrusion prevention and industrial monitoring. Despite the large number of proposed anomaly detection algorithms, many novel meth...

📖 Read original article


60. An entropic explanation of insistence on sameness in autism ​

Author: Przemys{\l}aw 'Sliwi'nski
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC, stat.ML

arXiv:2608.04616v1 Announce Type: new Abstract: An information theory-based framework is proposed in attempt to explain insistence on sameness in autism as an instance of a general behavior pattern in which an individual tries to reduce surprise and uncertainty. It offers a new definition of autism ...

📖 Read original article


61. Active Learning Guided Design Space Refinement for Scalable Multi-Objective Bayesian Optimization in Materials Discovery ​

Author: Alexandros Ntagiantas, Panagiotis Tsilimidos, George Giannakopoulos, Christoforos Rekatsinas, Panagiotis Krokidas
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci

arXiv:2608.04651v1 Announce Type: new Abstract: Advanced materials discovery increasingly relies on machine learning and Bayesian optimization to explore large discrete design spaces under limited evaluation budgets. However, conventional Bayesian optimization (BO) can become inefficient as candidat...

📖 Read original article


62. Differentiating Through Dual Prices: End-to-End Policy Learning Under Capacity Constraints ​

Author: Mohammadsaeed Haghi, Mahdi Salmani, Nima Kelidari
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04669v1 Announce Type: new Abstract: Many social services assign scarce resources, such as housing assistance or hospital interventions, to people who arrive one at a time: each arrival must receive a decision immediately, and the long-run usage of every resource must stay within its capa...

📖 Read original article


63. Diverse and Plausible Algorithmic Recourse via Tractable Recourse Distributions ​

Author: Anagha Sabu, Hrithik Suresh, Narayanan C. Krishnan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04677v1 Announce Type: new Abstract: Algorithmic recourse seeks to help individuals reverse unfavorable automated decisions by recommending actionable changes that achieve a desired outcome. As an individual usually has several distinct routes to a favorable decision, and different people...

📖 Read original article


64. The Sample Complexity of Distributionally Robust PAC Learning under Cressie--Read Divergences ​

Author: Elad Aigner-Horev, Daniel Rosenberg, Roi Weiss
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04686v1 Announce Type: new Abstract: We study distributionally robust PAC learning for the $0$--$1$-loss, where adversarial perturbations of the data distribution are constrained by a Cressie--Read divergence of order $k>1$ and radius $\rho\geq 0$. For hypothesis classes with VC dimension...

📖 Read original article


65. Personalized Federated Sparse Adaptation of Time-Series Foundation Models ​

Author: Priyanka Nihalchandani, Naman Srivastava, Varun Ojha, Pandarasamy Arjunan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.04695v1 Announce Type: new Abstract: Federated adaptation of time-series foundation models (TSFMs) is attractive for building energy forecasting because meter data are private, distributed, and highly non-IID. However, a single parameter-sharing strategy is unlikely to serve all pretraine...

📖 Read original article


66. Design Choices That Matter: A Functional ANOVA Analysis for Remote Sensing Multi-Label Classification ​

Author: Maryam Gholami Shiri, Eva Tuba, Sa\v{s}o D\v{z}eroski, Tome Eftimov, Ana Nikolikj
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.04702v1 Announce Type: new Abstract: Benchmarking deep learning (DL) models for multi-label classification (MLC) of remote sensing images (RSI) typically yields rankings that do not generalize beyond the evaluated datasets. In this work, we move beyond rankings by employing functional ana...

📖 Read original article


67. Benchmarking Deep Learning Models for Dense Event Classification of Offshore Wind Infrastructure in Sentinel-1 Time Series ​

Author: Thorsten Hoeser, Felix Bachofer, Claudia Kuenzer
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04706v1 Announce Type: new Abstract: Monitoring of offshore wind energy infrastructure life cycles, especially during the deployment phase, is an important contribution for stakeholders to make informed decisions in a phase of increasing deployment activities. ESA's Sentinel-1 Synthetic A...

📖 Read original article


68. A 6G Integrated Sensing and Communication Framework for Railway Intrusion Detection and Collision Prediction ​

Author: Ajeet Kumar Yadav, Sankaran Balasubramaniam, Aritra Chatterjee, Vinod Aduru, Yogesh Simmhan, Pandarasamy Arjunan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP

arXiv:2608.04710v1 Announce Type: new Abstract: Integrated Sensing and Communication (ISAC) combines sensing and communication to efficiently utilize wireless resources and is emerging as a key paradigm for next-generation wireless networks. By leveraging the wide bandwidth, high frequencies, and ma...

📖 Read original article


69. Attention, Anomalies! Handling Attention Layers in Unsupervised Federated Outlier Detection ​

Author: Mihailo Ili'c, Milo\v{s} Savi'c, Vladimir Kurbalija, Mirjana Ivanovi'c, Giancarlo Fortino, Du\v{s}an Jakoveti'c
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04753v1 Announce Type: new Abstract: Attention layers are the backbone of today's most powerful and impactful models. Models with multi-million and billion parameters rely on contextual knowledge provided by attention layers. However, their use goes well beyond just being the core compone...

📖 Read original article


70. IMFACT: Counterfactual Explanations for Time Series via Intrinsic Mode Function Substitution ​

Author: Udo Schlegel, Julian Rakuschek, Thomas Seidl, Andreas Holzinger, Tobias Schreck, Javier Del Ser
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04777v1 Announce Type: new Abstract: Oscillatory signals, such as vibration, carry class-discriminative information in specific frequency bands; perturbing them in raw feature space for counterfactual analysis easily destroys their temporal structure and produces physically implausible re...

📖 Read original article


71. Continual-Learning Physics-Informed Neural Networks for Parameterized Partial Differential Equations ​

Author: Xujia Chen, Xinyue Hu, Letian Chen, Yi Liu, Wenhui Fan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04778v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) incorporate governing equations into neural-network training and can approximate PDE solutions without requiring large observational datasets. Parameterized PINNs (ParamPINNs) further take physical parameters as...

📖 Read original article


72. Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation ​

Author: Yi Yang, Cong Qin, Xiaodan Liu, Chishui Chen, Qing Dong, Yan Zhang, Cao Liu, Zhao Yang, Lu Pan, Jiaye Lin, Yi Feng
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.04788v1 Announce Type: new Abstract: Large language model agents are commonly trained through reinforcement learning with sparse trajectory-level rewards, which offer limited guidance on how strongly individual tokens should be updated. On-Policy Self-Distillation (OPSD) addresses this by...

📖 Read original article


73. Above-ground Biomass Estimation with Geospatial Foundation Models ​

Author: Ghjulia Sialellia, Linus Scheibenreif, Jan Dirk Wegner, Konrad Schindler
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04792v1 Announce Type: new Abstract: Accurate estimation of Above-Ground Biomass (AGB) from satellite imagery is essential for the large-scale monitoring of carbon stocks, yet it remains a challenging regression task at global scale. Geospatial Foundation Models (GFMs) have recently emerg...

📖 Read original article


74. MGSB: Manifold Gated Signature Branch Pressure-Domain Baseline Architecture for Two-Phase Pipeline Flows Under Distributional Shift ​

Author: Issah Suleiman, Sormeh Serpoosh, Nadine Elkholy, Hicham Ferroudji, Mohammad Azizur Rahman, Matthew Hamilton
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04805v1 Announce Type: new Abstract: Leak detection models for multiphase pipelines often degrade when deployed under flow regimes that differ from training. Existing evaluations typically assess performance under in-distribution operating conditions, masking failures caused by regime tra...

📖 Read original article


75. Robust Control under Stationary Ambiguity ​

Author: Konrad J. Mueller, Amira Akkari, Ben Wood, Lukas Gonon
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP

arXiv:2608.04832v1 Announce Type: new Abstract: Control policies optimized in simulation can perform poorly in the real system when the parameters $x$ of the simulator are estimated from limited data but the resulting parameter uncertainty is not represented inside the simulation. A common way to in...

📖 Read original article


76. Training Crossroads for Recurrent Vision Transformers: Recurrence, Neural ODEs, and Deep Supervision ​

Author: Grzegorz Gruszczynski, Pawel Olszowiec, Michal Byra, Grzegorz Stefanski, Alberto Presta
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.04879v1 Announce Type: new Abstract: Vision Transformers (ViTs) achieve strong image-recognition performance, but their parameter count grows linearly with depth when each block is independently parameterized. Single-block recurrent ViTs (bViT) remove this growth by repeatedly applying on...

📖 Read original article


77. Variational Bounds for Perceptron Learning from Structured Data ​

Author: Francesco Camilli, Pierluigi Contucci, Federica Gerace, Emanuele Mingione
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, math-ph, math.MP, stat.ML

arXiv:2608.04882v1 Announce Type: new Abstract: We introduce a variational approach to a finite-temperature continuous-spin perceptron trained on a Gaussian mixture. The model allows for a broad class of concave utilities and log-concave separable prior measures on the spins. By combining the interp...

📖 Read original article


78. Consistency-Driven Co-Evolution for Self-Supervised Cross-Representation Learning ​

Author: Xuehang Guo, Pengyuan Li, Tom Hope, Tirthankar Ghosal, Manling Li, Qingyun Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.04926v1 Announce Type: new Abstract: As chart images, tabular data, and visualization code play increasingly important roles across diverse domains, cross-representation understanding across these modalities poses fundamental challenges for AI systems: the relationships across representat...

📖 Read original article


79. Optimal Training-Time Scaling in Gradual Adaptation ​

Author: Zonghuan Xu, Krishna Harish
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04927v1 Announce Type: new Abstract: In gradual adaptation, how should the training time on each task change as the number of intermediate tasks increases? We study this question for overparameterized linear regression tasks that change smoothly and share a zero-loss solution. With $N$ ta...

📖 Read original article


80. SVI-DAG: A Structured Variational Inference Approach to Bayesian Causal Discovery ​

Author: Shrenik Zinage
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04930v1 Announce Type: new Abstract: Bayesian causal discovery seeks to determine the posterior distribution of causal theories, which are interpreted as directed acyclic graphs (DAGs) that explain the observed data. The resulting posterior allows systematic reasoning regarding epistemic ...

📖 Read original article


81. CheMLFlow: An Open-Source Platform for Cheminformatics and Materials Informatics Applications ​

Author: Brendan Smith, Susana Lopez-Moreno, Eric Dolores-Cuenca, Sangil Kim, Jose L. Mendoza-Cortes, Nijamudheen Abdulrahiman
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cond-mat.other, cs.AI, physics.chem-ph

arXiv:2608.04942v1 Announce Type: new Abstract: CheMLFlow is an open-source platform for building and executing end-to-end, high-throughput, and agentic workflows for scientific and technological applications. CheMLFlow targets a common bottleneck in scientific machine learning development, where re...

📖 Read original article


82. SpecRoll: Fast-Slow Verifier-Feedback Adaptation for Speculative Reinforcement Learning Rollouts ​

Author: Nhat Minh Pham, Duy Tung Doan, Thi Duyen Ngo, Vinh Van Nguyen, Khac-Hoai Nam Bui
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.04962v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training improves the reasoning capabilities of large language models, but autoregressive rollout generation remains a major efficiency bottleneck. Speculative decoding can accelerate generation, yet applying it during ...

📖 Read original article


83. EvolveNet: Collaborative Harness Evolution for Agent Self-Improvement ​

Author: Jun Nie, Yonggang Zhang, Qianshu Cai, Yiu-ming Cheung, Xinmei Tian, Bo Han
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04968v1 Announce Type: new Abstract: The capabilities of an LLM agent depend not only on its model but on the harness: the executable program that constructs context, invokes tools, verifies results, and recovers from failure. Recent work shows that evolving the harness yields persistent ...

📖 Read original article


84. Stochastic Emulation using Generalized Stratified Sampling for Performance-Based Risk Optimization of Structures ​

Author: Isabela D. Rodrigues, Seymour M. J. Spence, Henrique M. Kroetz, Andr'e T. Beck
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.CO, stat.ML

arXiv:2608.05006v1 Announce Type: new Abstract: Metamodels are instrumental in reducing the computational burden associated with nested reliability analyses and optimization loops in Performance-Based Risk Optimization (PBRO) of structures under stochastic loads. In this context, stochastic emulator...

📖 Read original article


85. Canonical Joint Energy-Based Model on CIFAR-10: failure modes and practical indistinguishability of Predictor-Corrector and SGLD samplers ​

Author: Dmytro Knopov
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.05025v1 Announce Type: new Abstract: Joint Energy-Based Models (JEM) unify classification and generation within a single network and support out-of-distribution (OOD) detection. Canonical JEM training relies on stochastic gradient Langevin dynamics (SGLD); a theoretically motivated altern...

📖 Read original article


86. MultiPathFormer: Towards a Foundation Model for Multipath Wireless Propagation ​

Author: Blessed Guda, Kayley Sze, Carlee Joe-Wong
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP

arXiv:2608.05076v1 Announce Type: new Abstract: Recent advances in machine learning have enabled training of wireless foundation models, which aim to support tasks such as channel estimation, beam prediction, and localization based on wireless signals. Existing wireless foundation models typically p...

📖 Read original article


87. Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning ​

Author: Zheyuan Zhang, Manqing Mao, Hong Wang, Zhuoer Wang, Samson Koelle, Jie Yuan, Yanjun Lin, James Feng, Nikki Lijing Kuang, Yanfang Ye, Wei Niu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.05080v1 Announce Type: new Abstract: Critic-free group-based reinforcement learning has become a scalable approach for post-training large language models. However, most existing methods allocate the same number of rollouts to every task and trajectory state, even though some rollouts pro...

📖 Read original article


88. Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control ​

Author: Rohit Kumar Salla, Manoj Saravanan, Simon Stepputtis
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.05084v1 Announce Type: new Abstract: Diffusion policies are a powerful policy class for continuous control, but their iterative denoising process creates a substantial computational bottleneck. Reducing this cost requires adapting the number of denoising steps to the difficulty of each ac...

📖 Read original article


89. Capability-Gated Planning: Cost-to-Goal Discovery and the Limits of Myopic Experiment Selection ​

Author: Ahmed Hassoon, Mark Dredze
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.05085v1 Announce Type: new Abstract: Systems that automate scientific discovery must repeatedly decide which experiment to run, which hypothesis to test, which tool to build, and when to stop. Many systems make these decisions by maximizing a myopic score such as expected information gain...

📖 Read original article


90. MALT: Lightweight Curvature-Aware Muon via Diagonal Preconditioning ​

Author: Tongle Wu, Huanyu Dong, Ying Sun, Ziye Ma
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.05088v1 Announce Type: new Abstract: Muon has recently emerged as a promising alternative to AdamW for language model pretraining by orthogonalizing momentum matrices using Newton-Schulz iterations. Although Muon mitigates gradient anisotropy, it does not explicitly account for the curvat...

📖 Read original article


91. Multimodal Spatiotemporal Atmospheric Data Assimilation with Latent Flow-matching ​

Author: Dibyajyoti Chakraborty, Romit Maulik
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, physics.ao-ph, physics.flu-dyn

arXiv:2608.05103v1 Announce Type: new Abstract: Data assimilation (DA) uses Bayesian inference to update the state of a numerical forecast model with observed data. In this study, we propose a fundamentally different, unified approach to atmospheric data assimilation. We use latent video flow-matchi...

📖 Read original article


92. BnBERT-iPET: Sparse Few-Shot Language Modeling for Bengali via Lottery Ticket Pruning ​

Author: Sajib Hossain, Md Kamrus Samad, Anan Ghosh, Labib Imam Chowdhury, Nabeel Mohammed
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.05104v1 Announce Type: new Abstract: Deep neural networks have shown impressive success in NLP tasks owing to their complex structure and huge number of edges. Achieving state-of-the-art performance in natural language processing with a large pre-trained model such as BERT is expensive an...

📖 Read original article


93. Reward Structure Shapes the Interaction Between Episodic Exploration and Neural Memory in Reinforcement Learning ​

Author: Jai Malegaonkar, Rohan Patil, Henrik I. Christensen
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.05111v1 Announce Type: new Abstract: In partially observable reinforcement learning, agents face a dual bottleneck: they must explore to encounter rewarding states and retain that experience in memory to optimize their policies. Exploration bonuses and memory architectures are traditional...

📖 Read original article


94. DASyR-LLM: Domain-Aware Symbolic Regression with LLMs for Kinetic Model Discovery ​

Author: Roberto Aliaga Medina, Paulina Quintanilla, Antonio del Rio Chanona
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.SC

arXiv:2608.05120v1 Announce Type: new Abstract: Kinetic model discovery is a central challenge in chemical engineering, as accurate rate expressions are essential for understanding and controlling chemical and biological processes. Symbolic regression (SR) has emerged as a powerful data-driven appro...

📖 Read original article


95. SSTQ:Privacy-Preserving Vector Quantization via Subsampled Stochastic TurboQuant ​

Author: Adel Javanmard, David P. Woodruff, Vahab Mirrokni
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.05127v1 Announce Type: new Abstract: Achieving local differential privacy in distributed optimization while maintaining low communication cost remains challenging. Existing vector quantization methods, such as vqSGD, use high-dimensional geometric constructions but incur unfavorable dimen...

📖 Read original article


96. The Loss Does Not See the Basis, but Adam Does ​

Author: Devender Singh
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.05136v1 Announce Type: new Abstract: Gradient descent on a factored model $W = UV^\top$ is implicitly biased toward low-rank solutions, while Adam, starting from the same small initialization, is not. We trace the difference to the gauge symmetry of the loss, its invariance under $(U, V) ...

📖 Read original article


97. Leveraging Machine Learning to Gain Insights on Quantum Thermodynamic Entropy ​

Author: Srinivasa Rao. P
Published: 8/6/2026, 4:00:00 AM
Categories: quant-ph, cond-mat.stat-mech, cs.CL, cs.LG, physics.comp-ph

arXiv:2305.06177v1 Announce Type: cross Abstract: We present a thermodynamic analysis of a quantum engine that uses a single quantum particle as its working fluid, inspired by Szilard's classical single-particle engine. Our design is modeled after the classically-chaotic Szilard Map and involves a t...

📖 Read original article


98. AI-driven Multimodal Representation Learning for Latent Mediation Structure Discovery of Socioeconomic Disadvantage, Psychosocial Factors, and Cardiometabolic Multimorbidity: Insights from the All of Us Research Program ​

Author: Cong Cao, Shuangge Ma
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG, stat.AP

arXiv:2608.04016v1 Announce Type: cross Abstract: Social disadvantage is associated with multimorbidity, but the pathways linking social conditions to disease burden remain poorly understood. We developed an AI-driven multimodal mediation framework that integrates socioeconomic, psychosocial, clinic...

📖 Read original article


99. When More Becomes Less: Position-Dependent Repetition Effects in Language Models ​

Author: Han-yu Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.04021v1 Announce Type: cross Abstract: Cloze-style probes that vary how often a target token appears implicitly assume that more copies of a target affect prediction the same way regardless of where the readout slot sits. We show this assumption fails. Our two-probe design holds a repeate...

📖 Read original article


100. Monsoon Mayhem to Market Waves: Forecasting Fisheries Resilience in Sri Lanka ​

Author: Ruzaini Ahmed, Yohan Jayasinghe, Tharumini Gamage, Ifaz Ikram, Hasini Lawanya, Nirasha Munasinghe, Patalee Narasinghe, Nisansa de Silva, Sandareka Wickramanayake
Published: 8/6/2026, 4:00:00 AM
Categories: econ.GN, cs.LG, q-fin.EC

arXiv:2608.04023v1 Announce Type: cross Abstract: Sri Lanka's fisheries sector is important for jobs and food supply. Between 2019 and 2025, it faced several major problems at the same time, and how these events together affected fish production and prices is still not well understood. This study de...

📖 Read original article


101. A Multi-Cohort Validation of Censoring-Aware Conformal Lower Predictive Bounds for Pathology Survival Models ​

Author: Mingi Hong
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ME, cs.LG

arXiv:2608.04025v1 Announce Type: cross Abstract: Whole-slide survival models commonly provide risk rankings without calibrated statements about individual event times. We evaluate fixed-cutoff drcosarc, a post-hoc conformal wrapper for discrete-time multiple-instance learning survival heads using f...

📖 Read original article


102. NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts ​

Author: Mohammed I. Radaideh, Jeremy Moon, Andre Gala-Garza, Emma Son, Yug Shah, Majdi I. Radaideh
Published: 8/6/2026, 4:00:00 AM
Categories: cs.GR, cs.AI, cs.CV, cs.CY, cs.LG

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains largely unexplored. As an exmaple in nuclear engineering, general-purpose foundation models frequent...

📖 Read original article


103. The Cost of Binarizing Survival Outcomes in Clinical Prognostic Modeling ​

Author: Shashank Yadav, David M. Routman, Andrew Y. K. Foong
Published: 8/6/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG

arXiv:2608.04046v1 Announce Type: cross Abstract: Survival analysis is an established framework for analyzing time-to-event data, yet many clinical machine learning studies still binarize the outcome before model training. This practice excludes censored patients, collapses temporal information into...

📖 Read original article


104. Beyond the QBER Threshold: A Temporal QBER Based Machine Learning Framework for Multi Attack Detection in BB84 QKD ​

Author: Isha, Deepak Singh, Devesh Kumar, S. K Pal, Praful Hambarde, Amit Shukla
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.04047v1 Announce Type: cross Abstract: Conventional BB84 Quantum Key Distribution (QKD) systems rely on a fixed 11% Quantum Bit Error Rate (QBER) threshold to detect eavesdropping. However, stealthy attacks can remain below this threshold while still compromising channel security. This pa...

📖 Read original article


105. Statistical learning theory and Occam's razor: Regularization ​

Author: Tom F. Sterkenburg
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.04049v1 Announce Type: cross Abstract: The principle of Occam's razor, which instructs us to prefer simplicity in inductive inference, has attracted much scrutiny both in the philosophy of science and in machine learning. In either field, however, a justification for the principle has bee...

📖 Read original article


106. FM4WiFi: Flow Matching for Multi-AP Coordination in Dense Deployments of Beyond Wi-Fi 8 Networks ​

Author: Maksymilian Wojnar, Krzysztof Rusek, Katarzyna Kosek-Szott, Szymon Szott
Published: 8/6/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2608.04050v1 Announce Type: cross Abstract: Wi-Fi networks are moving beyond random channel access toward tightly coordinated operation across access points (APs), a shift reflected in Wi-Fi 8's multi-AP coordination (MAPC). However, the current MAPC specification restricts cooperation to AP p...

📖 Read original article


107. When Modalities Fail to Tango: Conformal Backdoor Detection in Multimodal Contrastive Learning ​

Author: Yiming Chen, Kemou Li, Haiwei Wu, Jiantao Zhou
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.CV, cs.LG

arXiv:2608.04052v1 Announce Type: cross Abstract: Backdoor attacks in multimodal contrastive learning (MCL) have garnered growing attention in recent years, as many downstream tasks critically depend on pre-trained MCL models. Existing detection-based defenses predominantly rely on the CLIPScore met...

📖 Read original article


108. Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding ​

Author: Mohnish Raj, Suraj Kumar, Soumi Chattopadhayay, Chandranath Adak, Ayan Dutta
Published: 8/6/2026, 4:00:00 AM
Categories: cs.MM, cs.AI, cs.CL, cs.LG

arXiv:2608.04054v1 Announce Type: cross Abstract: Multimodal intent recognition requires understanding not only what textual, acoustic, and visual signals share, but also how they disagree. Such disagreement is frequently class-informative; for example, lexical positivity accompanied by incongruent ...

📖 Read original article


109. Unifying quantum measurement constructions via a relative-entropy minimum change principle ​

Author: Nana Liu, Mark M. Wilde
Published: 8/6/2026, 4:00:00 AM
Categories: quant-ph, cond-mat.stat-mech, cs.IT, cs.LG, math.IT

arXiv:2608.04055v1 Announce Type: cross Abstract: The minimum change principle provides an information-theoretic characterization of the Bayes reversal channel in classical probability theory and has recently been proposed as a framework for extending Bayes' rule to quantum information theory. Using...

📖 Read original article


110. Learning Sexism Detection Using Multi-Agent Perspectivist Preference Optimization ​

Author: Hadi Mohammadi, Tina Shahedi, Robert A. Bagheri, Mehdi Dastani, Masoume M. Raeissi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.CY, cs.LG

arXiv:2608.04056v1 Announce Type: cross Abstract: When people label text for sexism, they often disagree, and not because some of them are wrong: they genuinely perceive sexism differently. Most NLP systems discard this disagreement by collapsing it into a majority vote. We propose the Multi-Agent P...

📖 Read original article


111. FBID: Adaptive Personalized Federated Learning for Robust Out-of-Distribution Attack Detection in IoT Networks ​

Author: An Khanh Bui, Cong Thanh Nguyen, Hoang-Anh Pham, Hoang Thai Dinh, Diep N. Nguyen
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.04073v1 Announce Type: cross Abstract: Personalized Federated Learning (PFL) has emerged as a promising solution for intrusion detection in heterogeneous IoT environments, as it can improve local adaptation under highly Non-Independent and Identically Distributed (non-IID) data distributi...

📖 Read original article


112. InvFlowFD: Reference-Free and Background-Set-Free Perceptual Music Quality Metric with Flow Matching Inversion ​

Author: Alon Ziv, Harel Pogoda, Yossi Adi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG

arXiv:2608.04142v1 Announce Type: cross Abstract: Existing reference-free methods for evaluating music perceptual quality alleviate the need for paired noisy-clean data, but they still rely on a background set, which is used to compute aggregated statistics of clean audio samples. In this work, we p...

📖 Read original article


113. Neighborhood-Aware Dual Biomedical Entity Linking ​

Author: Yicheng Tao, Jie Liu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.IR, cs.CL, cs.LG

arXiv:2608.04144v1 Announce Type: cross Abstract: Biomedical entity linking grounds mentions in clinical and scientific text to entities in a curated knowledge base (KB) with ontological structure, which supports downstream applications such as literature-scale information extraction and patient-rec...

📖 Read original article


114. Sublogarithmic Swap Regret in Multiplayer General-Sum Games via Hybrid Regularization ​

Author: Taira Tsuchiya
Published: 8/6/2026, 4:00:00 AM
Categories: cs.GT, cs.LG

arXiv:2608.04149v1 Announce Type: cross Abstract: Swap regret governs the rate at which uncoupled learning dynamics converge to correlated equilibria in multiplayer general-sum games. Under full-information feedback, the best previous guarantee when every player follows the same dynamics grows logar...

📖 Read original article


115. BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding ​

Author: Yangxuan Zhou, Sha Zhao, Yuning Chen, Chen Wu, Jiquan Wang, Shijian Li, Gang Pan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.04156v1 Announce Type: cross Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it requires workflows connecting natural-language instructions, signal processing, quantitative evidence, and scientific interpretation. We term this capa...

📖 Read original article


116. Mind the Cap: Output-Budget Regimes Change the Measured Multilingual Reasoning Gap ​

Author: Ankit Goyal, Jaideep Ray
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.04160v1 Announce Type: cross Abstract: Multilingual evaluations report accuracy at a single output-token cap, but languages need different numbers of tokens to express the same content, so the cap is a hidden experimental variable. We test whether the native-vs-translate gap on MGSM (Germ...

📖 Read original article


117. Adversarially Robust Abductive Fusion of Pre-trained Transformer-based Perception Models ​

Author: Mario Leiva, Yue Ma, Qinru Qiu, Gerardo Simari, Paulo Shakarian
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG, cs.LO

arXiv:2608.04190v1 Announce Type: cross Abstract: Deploying pre-trained perception models in novel environments degrades their accuracy under distributional shift, and assembling them alone does not recover it: combiners such as majority voting trade recall for precision and are brittle to coordinat...

📖 Read original article


118. SiMDex: Mining Similar Egocentric Videos for Cross-Embodiment Dexterous Manipulation ​

Author: Nie Lin, Takehiko Ohkawa, Sijin Chen, Ruoshi Wen, Zhuohang Li, Liqun Huang, Zhengming Zhu, Yiming Bao, Yunfei Li, Minjie Cai, Xiao Ma, Wei Xu, Yoichi Sato
Published: 8/6/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2608.04196v1 Announce Type: cross Abstract: Recent years have witnessed an explosive trend of scaling ego-centric human videos for robot manipulation, yet it remains unclear which data actually benefits dexterous manipulation. We present SiMDex, a similarity-based data mining framework that ca...

📖 Read original article


119. From Financial Sentiment Classification to Return Predictability: A QLoRA Benchmark of Large Language Models ​

Author: Fusheng Luo
Published: 8/6/2026, 4:00:00 AM
Categories: q-fin.MF, cs.LG

arXiv:2608.04200v1 Announce Type: cross Abstract: Financial sentiment classifiers are commonly evaluated against human labels, but strong linguistic performance does not necessarily imply economically useful return predictability. This study separates these questions through two experiments. First, ...

📖 Read original article


120. TIDE: A Physically Diverse 3D Turbulence Benchmark Dataset for Advancing Scientific Machine Learning ​

Author: Yilong Dai, Yiming Sun, Yiheng Chen, Shengyu Chen, Peyman Givi, Xiaowei Jia, Runlong Yu
Published: 8/6/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG

arXiv:2608.04222v1 Announce Type: cross Abstract: Turbulence is a central testbed for machine learning on physical dynamics because its governing laws are known exactly. However, most existing studies remain in 2D, while 3D turbulence has fundamentally different physics and is far more costly to sim...

📖 Read original article


121. Multimodal Alignment Through Joint Kernel Entropic Gromov--Wasserstein Optimal Transport ​

Author: Yixuan Florence Wu, Yilun Zhu, Naichen Shi
Published: 8/6/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.ML, stat.TH

arXiv:2608.04234v1 Announce Type: cross Abstract: We study the problem of aligning data from multiple modalities into a shared representation space, focusing on settings where strong pretrained unimodal encoders are available but cross-modal paired data are scarce. We propose a structure-preserving ...

📖 Read original article


122. Dynamical Lie Algebras Cannot Describe Shallow QAOA: Cragged Terrains, Barren Plateaus, and Empirical Hardness Models ​

Author: Harrison Copp, Charlton Li, An\v{z}ej Margeta-Cacace, Amy Qiao
Published: 8/6/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG

arXiv:2608.04252v1 Announce Type: cross Abstract: The dynamical Lie algebraic (DLA) theory of variational quantum algorithms (VQAs) predicts commonplace exponentially vanishing loss and gradient variances for sufficiently deep parametrized circuits. In this work, we show that these predictions fail ...

📖 Read original article


123. PriDyG: Privacy-preserving Dynamic Graph Inference with LLM-GNN Collaboration ​

Author: Yuyang Xia, Ruixuan Liu, Li Xiong
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.04255v1 Announce Type: cross Abstract: Graph inference over relational data can expose sensitive edge information, and this risk becomes more severe in dynamic graphs, where repeated model updates cause privacy loss to accumulate. We formulate Edge-level Differentially Private Dynamic Gra...

📖 Read original article


124. The RAIL Principles for Neurosymbolic AI: Reasoning, Assurances, Interfacing and Learning ​

Author: Agnese Chiatti, Michael Cochez, Cristina Cornelio, Sebastijan Dumancic, Artur d'Avila Garcez, Luis C. Lamb, Lia Morra, Mathias Niepert, Robert Peharz, Alberto Speranzon, Maarten Stol, Annette Ten Teije, Thiviyan Thanapalasingam, Frank Van Harmelen, Emile Van Krieken, Antonio Vergari, Benjie Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.04285v1 Announce Type: cross Abstract: Neurosymbolic AI systems that integrate machine learning and symbolic reasoning are rapidly gaining attention. They complement the data-intensive statistical approaches of neural networks and language models with symbolic reasoning algorithms to func...

📖 Read original article


125. Eliciting Intrinsic Hallucinations in LLMs via Semantically Equivalent Adversarial Attacks ​

Author: Atri Vivek Sharma, Brian Formento, Alessio Lomuscio
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.04286v1 Announce Type: cross Abstract: Large language models (LLMs) are often used in conjunction with external knowledge sources to improve their factual accuracy and decrease hallucinations, through methods such as Retrieval-Augmented Generation (RAG). However, these systems remain susc...

📖 Read original article


126. Trident : How to Break Deep Reinforcement Learning Cyber Defenses (Agentic) ​

Author: Ryozo Masukawa, Ian Bryant, Armita Kazeminajafabadi, Sanggeon Yun, Hyunwoo Oh, SungHeon Jeong, Nathaniel D. Bastian, Mahdi Imani, Mohsen Imani
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG, cs.MA

arXiv:2608.04317v1 Announce Type: cross Abstract: Autonomous cyber defense systems based on Deep Reinforcement Learning (DRL) have attracted significant research attention, yet remain evaluated almost exclusively against static, heuristic red agents, leaving their robustness against adaptive threats...

📖 Read original article


127. Right Reset: Chunking by Prefix Removal ​

Author: Mike Vegeto
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.04330v1 Announce Type: cross Abstract: Removing the left context from a causal language model reveals a useful kind of boundary: an edge where the model processes the same right-hand tokens with little change. We turn this observation into prefix-removal probing and introduce Right Reset ...

📖 Read original article


128. iStructTab: Structured Feature Sequencing for Multimodal Learning of Image and Tabular Data ​

Author: Al Zadid Sultan Bin Habib, Md Younus Ahamed, Prashnna Gyawali, Gianfranco Doretto, Donald A. Adjeroh
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, stat.ML

arXiv:2608.04348v1 Announce Type: cross Abstract: Multimodal learning of images and tabular data is often impaired by ineffective representations, resulting in redundancy, dispersion, and generalization problems. To tackle this challenge, we introduce Graph-Enhanced Descriptor Sequencing (GEDS), a s...

📖 Read original article


129. NeuMoSync: End-to-End Neuromodulatory Control for Plasticity and Adaptability in Continual Learning ​

Author: Seyed Roozbeh Razavi Rohani, Khashayar Khajavi, Wesley Chung, Mandana Samiei, Mo Chen
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.04358v1 Announce Type: cross Abstract: Continual learning (CL) requires models to learn tasks sequentially, yet deep neural networks often suffer from plasticity loss and poor knowledge transfer, which can impede their long-term adaptability. Drawing high-level inspiration from global neu...

📖 Read original article


130. Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation ​

Author: Scott H. Hawley
Published: 8/6/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS

arXiv:2608.04378v1 Announce Type: cross Abstract: Collaborative music agents need internal representations rich enough to support both understanding and generation, yet flexible enough for a workflow where the human retains agency. We present a hierarchical self-supervised ``world model'' for symbol...

📖 Read original article


131. Deltoris: Enabling Real-time VLA Inference in Embodied AI via Bit-level Sparsity and Speculative Inference ​

Author: Zheng Liu, Zeyu Guo, Zihan Liu, Anbang Wu, Han Zhao, Fangxin Liu, Zhezhi He, Yinhe Han, Jingwen Leng, Minyi Guo, Yiming Gan, Yu Feng
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AR, cs.LG

arXiv:2608.04428v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have emerged as a key component in embodied AI. Among existing approaches, diffusion-based VLA models achieve superior motion quality and generalization. However, diffusion-based VLA models are compute-intensive an...

📖 Read original article


132. The Price of Isolation: Estimating the Ecosystem Cost of Symmetric Two-Sided A/B Testing ​

Author: Yuanyuan Shen, Yiren Yan, Wenjie Li, Chunhui Zhu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, cs.SI, stat.ME

arXiv:2608.04432v1 Announce Type: cross Abstract: On two-sided content platforms, symmetric two-sided isolation (assigning matched fractions of creators and viewers to isolated treatment and control submarkets) is widely used for creator-side and cold-start experiments because it removes cross-arm m...

📖 Read original article


133. A Counterexample to Fourier Alignment in Single-Neuron Modular Addition ​

Author: Gautam Neelakantan Memana
Published: 8/6/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.NE

arXiv:2608.04451v1 Announce Type: cross Abstract: We give a negative solution to MAIS-O60. We first construct an example in which an initially active ReLU neuron becomes completely inactive in finite time and thereafter remains frozen at a limit whose Fourier energy is equally distributed among all ...

📖 Read original article


134. Beyond Global Routing Aggregation: Phase-Aware Expert Merging for MoE Vision-Language Models ​

Author: Hongyu Zhang, Cheng Yan, Xiang Xia, Wuyang Zhang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.04454v1 Announce Type: cross Abstract: Mixture-of-experts vision-language models (MoE-VLMs) increase model capacity with sparse expert activation, yet deployment requires storing the full expert pool. Training-free expert merging reduces this burden, and many routing-based methods aggrega...

📖 Read original article


135. Multi-Objective Ranking for Live-Streaming: Balancing Fresh and Delayed Signals with Segment-Aware Targeting ​

Author: Xiaoyi Gu, Julia Tavares, Eder Santana, Carlos Mendoza-Cardenas, Nikita Mishra, Saad Ali
Published: 8/6/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.04455v1 Announce Type: cross Abstract: One of the most challenging problems entertainment live-streaming services face in recommendation systems is that user behaviors are sparse and delayed, and interaction data exhibits bias for different user segments. Unlike e-commerce applications wh...

📖 Read original article


136. DeepInvert: Semi-Supervised Embedding Inversion Against Obfuscated Language Models ​

Author: Zhicong Huang, Cheng Hong, Tao Wei
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG

arXiv:2608.04477v1 Announce Type: cross Abstract: Cloud-based language model services routinely process prompts containing sensitive information. Obfuscation-based defenses---including ObfusLM, SentinelLMs, TextObfuscator, and DPNR---mitigate this risk by transforming prompt representations before t...

📖 Read original article


137. DIVE: Dynamic Iterative Visual Evidence Construction for Efficient Vision-Language Models ​

Author: Chen Zhong, Xiao An, Zijie Wang, Jiepan Li, Guangyi Yang, Wei He
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.04496v1 Announce Type: cross Abstract: Visual inputs in vision-language models (VLMs) are often encoded into substantially longer token sequences than text, making visual tokens a major bottleneck for efficient inference. Abundant recent methods address this bottleneck by scoring token im...

📖 Read original article


138. ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance ​

Author: Javier Rodriguez-Juan, Hiba Arnaout, Jose Garcia-Rodriguez, David Tom'as, Iryna Gurevych
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.04524v1 Announce Type: cross Abstract: Synthetic generation of Cognitive Behavioral Therapy (CBT) sessions is challenged by two competing demands: adhering to strict therapeutic structure while modeling the resistant, unpredictable behavior of real patients. Existing script-based methods ...

📖 Read original article


139. EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive Decision Tasks ​

Author: Pau Arnal, Khaled Denfir, Danylo Smahliuk, Amrut Avhad, Marcus A. Castro
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.04549v1 Announce Type: cross Abstract: Frontier LLMs are increasingly put to use on open-ended complex questions, different in nature from the ones they are typically evaluated on. We dedicate more than 4,000 human expert hours to evaluate a selection of six frontier LLMs on a member of t...

📖 Read original article


140. Relational Response Fields: A General Theory of Black-Box LLM Response Consistency and Recovery ​

Author: Song Zichen
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.04552v1 Announce Type: cross Abstract: Black-box language-model reliability is commonly pursued by sampling, prompting, voting, verifying, or iteratively revising individual answers. We ask a prior question: \emph{what determines whether a collection of black-box responses is recoverable ...

📖 Read original article


141. Relevant but Incomplete: Referential Dangling as a Paradigm-Level Failure Mode in Hard Prompt Compression ​

Author: Zhengpei Hu, Kai Li, Dapeng Fu, Xuechao Zou, Yuanhao Tang, Yue Li, Tengfei Cao, Jianqiang Huang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.04569v1 Announce Type: cross Abstract: Hard prompt compression reduces long-context inference cost by independently scoring tokens, sentences, or chunks and retaining the highest-scoring units under a budget. We identify a structural failure in this procedure: independent selection can sp...

📖 Read original article


142. On MUON optimization: From non-convergence to an error analysis with Polar Express and the Newton-Schulz polynomial from implementations ​

Author: Thang Do, Steffen Dereich, Arnulf Jentzen
Published: 8/6/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2608.04607v1 Announce Type: cross Abstract: Stochastic gradient descent (SGD) optimization methods are the standard instruments for the training of deep neural networks (DNNs). In many relevant artificial intelligence (AI) systems - such as popular large language models (LLMs)-not the standard...

📖 Read original article


143. Automatic Statistical Test for Rationally Expressible Algorithms by Selective Inference, with Applications to Feature Selection ​

Author: Teruyuki Katsuoka, Tomohiro Shiraishi, Shuichi Nishino, Ichiro Takeuchi
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.04667v1 Announce Type: cross Abstract: Selective inference (SI) provides statistically valid $p$-values for hypotheses selected by applying an algorithm to the data, correcting for the bias that arises when the same data are used both to select and to test a hypothesis. Developing an SI p...

📖 Read original article


144. Kathleen Writes: Autoregressive Generation and Data Scaling Without Attention ​

Author: George Fountzoulas
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.04678v1 Announce Type: cross Abstract: Papers 1-2 of the Kathleen series showed that a byte-level, attention-free architecture built from a wavetable encoder and multi-scale reverberant state can match strong baselines on classification at ~450-700K parameters, without pretraining. We ask...

📖 Read original article


145. Suppression Sticks, Locality Is Fragile: A Closed-Loop Target-and-Control Audit of Task-Vector Negation in VLA Policies ​

Author: Shaoguang Wang, Weiyu Guo, Rushi Dai, Yiren Zhao, Yandong Guo, Hui Xiong
Published: 8/6/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.04692v1 Announce Type: cross Abstract: Task-vector arithmetic offers a closed-form way to modify a model, yet its behavioral locality remains unclear in closed-loop robot control. We present a target-and-control audit of per-skill task-vector subtraction from multitask vision-language-act...

📖 Read original article


146. What We Observe as LLM Behavior Can Be a Side-effect of Inference Backend ​

Author: Shahed Masoudian, Passant Shafaei, Monorama Swain, Markus Schedl
Published: 8/6/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2608.04714v1 Announce Type: cross Abstract: Benchmark scores are reported as properties of a model, yet the inference framework used to produce them, such as HuggingFace, vLLM, or Ollama, are considered non-influential and their names and versions are almost never disclosed. In this work we in...

📖 Read original article


147. Privileged, but Biased: How PI-Conditioned Teachers Break Self-Distillation ​

Author: Sarthak Harne, Chinmay Karkar, Yash Pandya, Ahmed Awadallah, Akshay Nambi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.04794v1 Announce Type: cross Abstract: Self-distillation (SD) has emerged as a compute-efficient alternative to reinforcement learning with verifiable rewards: a self-teacher, conditioned on privileged information (PI) about the answer such as a reference solution, supplies dense per-toke...

📖 Read original article


148. Intrinsic-Hybrid Latent Diffusion Models for Generative Modeling on Unknown Manifolds ​

Author: Yizhu Wang, Mu Niu, Xiaochen Yang
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.04827v1 Announce Type: cross Abstract: We introduce the Intrinsic Hybrid Latent Diffusion Model (ILDM), a generative framework that integrates probabilistic dimensionality reduction with geometry-aware diffusion on unknown manifolds. While diffusion models (DMs) have achieved state-of-the...

📖 Read original article


149. Nonparametric Goodness-of-fit Testing under Covariate Shift ​

Author: Zhen Hou, Dong Xia
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, math.ST, stat.ML, stat.TH

arXiv:2608.04860v1 Announce Type: cross Abstract: This paper develops procedures for nonparametric goodness-of-fit testing under covariate shift, where labelled data are drawn from a source population but goodness-of-fit is evaluated for a target population. The distribution mismatch is quantified b...

📖 Read original article


150. The Neural Echo: A Signal Processing Perspective for Understanding Neural Networks ​

Author: Chongbiao Wang, Daniel Gaa, Joachim Weickert, Karl Schrader
Published: 8/6/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.04864v1 Announce Type: cross Abstract: We introduce the neural echo as a tool for understanding the behavior of neural networks. It generalizes the model-based concepts of impulse responses, diffusion echoes, and filter echoes to learning-based methods. It provides local, space-adaptive i...

📖 Read original article


151. A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination ​

Author: Wenxiao Zhao, Dong Liu, Kaiyi Xu, Feng Liu, Zhen Zhao, Fei Ben, Shu Wang, Wenhao Li, Yingnian Wu, Fenghua Ling, Haobo Li, Lei Bai
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.04872v1 Announce Type: cross Abstract: Symbolic regression aims to discover closed-form equations from data, but existing LLM-guided methods often rely on a unified proposal loop that compresses heterogeneous search failures into a scalar score and a single prompt. We propose A-SR, a self...

📖 Read original article


152. When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs ​

Author: Jiaming Cheng, Subhransu Das, Rajiv Ramnath
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.04893v1 Announce Type: cross Abstract: Multi-agent LLM systems relay key--value caches instead of text and credit their gains to exchanged ``latent thoughts''. That credit is a claim about \emph{which} example's cache is relayed, not merely that one is. We audit it causally in released sy...

📖 Read original article


153. Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation ​

Author: Zehua Chen, Junyou Wang, Yuxuan Jiang, Zhenying Fang, Yusheng Dai, Jianfei Chen, Ziwei Liu, Jun Zhu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.SD

arXiv:2608.04902v1 Announce Type: cross Abstract: Video-to-audio (V2A) generation extends image-to-audio generation (I2A) by introducing consecutive frames that provide essential temporal cues for audio synthesis. However, existing conditional diffusion-based V2A methods typically enhance visual con...

📖 Read original article


154. State2State: Environment-Derived Mid-Training for LLM Agents ​

Author: Xuanyu Lei, Yiqi Zhu, Chenliang Li, Kaiming Liu, Peng Li, Ming Yan, Jieping Ye, Ya-Qin Zhang, Yang Liu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.04934v1 Announce Type: cross Abstract: Training LLM agents commonly relies on supervised fine-tuning from expert trajectories or online reinforcement learning over human-specified tasks with handcrafted verifiers. Though effective, both remain bottlenecked by externally specified tasks an...

📖 Read original article


155. A geometry-based deep equilibrium model for image restoration under multiplicative Gamma noise ​

Author: Shengkun Yang, Luca Ratti, Zhichang Guo
Published: 8/6/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA

arXiv:2608.04944v1 Announce Type: cross Abstract: We propose a deep learning framework for image restoration from images degraded by both multiplicative Gamma noise and blur. Unlike conventional deep equilibrium (DEQ) models that rely on implicit neural regularization, the proposed method learns an ...

📖 Read original article


156. WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models ​

Author: Bohai Gu, Yueyang Yuan, Taiyi Wu, Dazhao Du, Jian Liu, Xiaoyi Pang, Jie Zhang, Xiaocheng Lu, Haobin Zhong, Xiaotong Zhao, Alan Zhao, Song Guo
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.04964v1 Announce Type: cross Abstract: Interactive video world models are essential for long-horizon planning and exploration, yet they suffer from compounding errors. Post-training methods such as reinforcement learning (RL) can improve these models, but they hit a verification bottlenec...

📖 Read original article


157. Protoreasoning in Tiny Transformers ​

Author: Eduardo Valle, Fergal Reid
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.04980v1 Announce Type: cross Abstract: We show that tiny transformers can profitably employ a simple form of Chain of Thought, which we call protoreasoning, allowing us to study step-by-step reasoning on ~1M-parameter models and opening up opportunities for much more detailed experimentat...

📖 Read original article


158. Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes ​

Author: Junlin Han, Shengbang Tong, David Fan, Minghao Chen, Philip Torr, Filippos Kokkinos, Mike Lewis
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.MM

arXiv:2608.05000v1 Announce Type: cross Abstract: Vision offers a critical axis for advancing foundation models, driving a shift towards natively unified multimodal pretraining. Despite this momentum, the design space and the fundamental mechanisms of how modalities interact during unified training ...

📖 Read original article


159. OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents ​

Author: Jingsheng Zheng, Xinyuan Fang, Jintian Zhang, Zhengke Gui, Huajun Chen, Ningyu Zhang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC, cs.LG, cs.MA

arXiv:2608.05013v1 Announce Type: cross Abstract: LLM agents are increasingly applied to open-ended everyday requests that span work, study, and life. These tasks are long-horizon, cross-environment, and multimodal, forcing the agent to preserve goals and constraints across many steps while navigati...

📖 Read original article


160. Revealed Rationality: Label-Free Evaluation and Regularization from Representation Theorems ​

Author: Isaiah Andrews
Published: 8/6/2026, 4:00:00 AM
Categories: econ.TH, cs.AI, cs.LG

arXiv:2608.05015v1 Announce Type: cross Abstract: Representation theorems in decision theory establish that behavior satisfies certain axioms if and only if it can be rationalized by a well-defined objective. I argue that this ``if and only if'' structure provides a potentially useful foundation for...

📖 Read original article


Author: Zidu Yin, Yuankai Qi, Dong Gong, Ehsan Abbasnejad, Kun Yue, Javen Qinfeng Shi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.SI, cs.LG

arXiv:2608.05016v1 Announce Type: cross Abstract: Predicting the existence and type of links (edges) between nodes in a multi-relational graph is key for applications from social interaction prediction to knowledge relationship identification. Enhancing local features with relevant global informatio...

📖 Read original article


162. Short-term load forecasting under EU-AI Act Requirements in Safety-Critical Environments: Results from a 41-day live challenge on the aggregated German transmission-grid load ​

Author: Thomas Bartz-Beielstein
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.05018v1 Announce Type: cross Abstract: Short-term load forecasting (STLF) play a vital role in the electric power industry. It serves infrastructure that European and German law designate as critical. Determinism, reproducibility, and auditability are engineering requirements rather than ...

📖 Read original article


163. SparseDitto: Customizing GPU Kernels for Different Sparsity Patterns with LLM-Based Agentic System ​

Author: Shiyang Li, Guangyan Sun, Jinwei Tang, Yanzhi Wang, Mingyi Hong, Caiwen Ding
Published: 8/6/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.05033v1 Announce Type: cross Abstract: Sparse matrix kernels are fundamental to scientific computing, graph analytics, and machine learning. Their GPU performance depends strongly on the input sparsity pattern and execution strategy. For the same SpMM on the same matrix, cuSPARSE exhibits...

📖 Read original article


164. MarsCast: Transfer Learning of AI Weather Foundation Models to Planetary Atmospheres ​

Author: M. L. Carroll, J. Li, S. D. Guzewich, G. Villanueva, J. A. Caraballo-Vega, M. J. Frost
Published: 8/6/2026, 4:00:00 AM
Categories: astro-ph.EP, cs.AI, cs.CV, cs.LG

arXiv:2608.05054v1 Announce Type: cross Abstract: We investigate the transferability of Earth weather foundation models to planetary atmospheres by adapting the GraphCast graph neural weather forecasting model to Mars. While GraphCast achieves state-of-the-art performance for terrestrial forecasting...

📖 Read original article


165. Provable Limits and Certified Deferral for Verbalized Uncertainty in Small Language Models ​

Author: Jianru Shen
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.05064v1 Announce Type: cross Abstract: Small open-weight language models increasingly run in private, offline, and cost-sensitive settings, where the key deployment question is not only what a model answers but when it should defer to a human. We study whether verbalized confidence can su...

📖 Read original article


166. Representational separation between unitary and channel quantum generative models via shared classical randomness at shallow depth ​

Author: Arunava Majumder, Marius Krumm, Hendrik Poulsen Nautrup, Hans J. Briegel
Published: 8/6/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG, stat.ML

arXiv:2608.05110v1 Announce Type: cross Abstract: Near-term quantum hardware limits circuit depth and often imposes geometrically local connectivity for quantum generative models, restricting the output distributions accessible to shallow unitary Born models. Introducing stochasticity into a unitary...

📖 Read original article


167. Stable Density Ridges: Consistency and Convergence of Subspace Constrained Mean Shift ​

Author: Wanli Qiao
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.05112v1 Announce Type: cross Abstract: The Subspace Constrained Mean Shift (SCMS) algorithm is a popular nonparametric method for extracting density ridges, which serve as a low-dimensional representation of high-dimensional data. It is a widely held belief in the literature that SCMS tra...

📖 Read original article


168. Robust and Efficient Motion Reasoning for Privacy-Aware Classroom Incident Recognition ​

Author: Paritosh Parmar, Landy Lan, Hong Yang, Chen Yi, Chiat Pin Tay
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.ET, cs.HC, cs.LG

arXiv:2608.05115v1 Announce Type: cross Abstract: Can computer vision help make classrooms safer? In this pilot study, we investigate privacy-aware and computationally efficient classroom incident recognition from CCTV-style observations. This setting remains underexplored, with limited benchmarks a...

📖 Read original article


169. Chained Recursive Language Models for Multi-Iteration Reasoning ​

Author: Purbesh Mitra, Sennur Ulukus
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.IT, cs.LG, eess.SP, math.IT

arXiv:2608.05124v1 Announce Type: cross Abstract: Long context reasoning in large language models (LLMs) is usually constrained by the fact that a single inference trajectory has to simultaneously explore the context, store intermediate state, verify evidence, and produce the final answer. This beco...

📖 Read original article


170. Predicting Brain Morphometry with MT-GNN: Mesh Evolution in Continuous Time with Graph-Based Metric Tensor Embeddings ​

Author: Hao Ding, Daniel Semchin, Paul M. Thompson, Boris Gutman
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.05132v1 Announce Type: cross Abstract: Predicting how a subcortical structure's shape will evolve from a few prior scans could support prognosis and clinical-trial enrichment. Existing longitudinal mesh predictors either extrapolate shape trajectories via high-dimensional embeddings or re...

📖 Read original article


171. Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning ​

Author: Yinghui He, Ling Yang, Jiarui Liu, Yongjin Yang, Lechen Zhang, Yingcheng Wu, Zhenfei Yin, Mengdi Wang, Sanjeev Arora
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.05139v1 Announce Type: cross Abstract: Long-horizon reasoning in recent LLMs demands that the model switch between distinct skills inside a reasoning chain, such as first doing a math derivation, then using the result to plan a schedule. We call such problems cross-skill long-horizon task...

📖 Read original article


172. OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling ​

Author: Indraneil Paul, Falko Helm, Goran Glava\v{s}, Iryna Gurevych
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE

arXiv:2608.05141v1 Announce Type: cross Abstract: Context lengths of language models (LMs) have dramatically increased, driven by the demands for in-context learning, self-improvement, and long-horizon agentic workflows. Existing long-context corpora, however, are dominated by books, academic articl...

📖 Read original article


173. GFlowNet Training by Policy Gradients ​

Author: Puhua Niu, Shili Wu, Mingzhou Fan, Xiaoning Qian
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2408.05885v3 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) have been shown effective to generate combinatorial objects with desired properties. We here propose a new GFlowNet training framework, with policy-dependent rewards, that bridges keeping flow balance of GFlowNe...

📖 Read original article


174. Unforgettable Generalization in Language Models ​

Author: Eric Zhang, Leshem Choshen, Jacob Andreas
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2409.02228v2 Announce Type: replace Abstract: When language models (LMs) are trained to forget (or "unlearn'') a skill, how precisely does their behavior change? We study the behavior of transformer LMs in which tasks have been forgotten via fine-tuning on randomized labels. Such LMs learn to ...

📖 Read original article


175. Regularization can make diffusion models more efficient ​

Author: Mahsa Taheri, Johannes Lederer
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH

arXiv:2502.09151v3 Announce Type: replace Abstract: Diffusion models are one of the key architectures of generative AI. Their main drawback, however, is the computational costs. This study indicates that the concept of sparsity, well known especially in statistics, can provide a pathway to more effi...

📖 Read original article


176. Towards Understanding Gradient Flow Dynamics of Homogeneous Neural Networks Beyond the Origin ​

Author: Akshay Kumar, Jarvis Haupt
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2502.15952v3 Announce Type: replace Abstract: Recent works exploring the training dynamics of homogeneous neural network weights under gradient flow with small initialization have established that in the early stages of training, the weights remain small and near the origin, but converge in di...

📖 Read original article


177. GenAI-Powered Inference ​

Author: Kosuke Imai, Kentaro Nakamura
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2507.03897v3 Announce Type: replace Abstract: We introduce GenAI-Powered Inference (GPI), a statistical framework for both causal and predictive inference using unstructured data, including text and images. GPI leverages open-source Generative Artificial Intelligence (GenAI) models---such as l...

📖 Read original article


178. Uncertainty-aware Predict-Then-Optimize Framework for Equitable Post-Disaster Power Restoration ​

Author: Lin Jiang, Dahai Yu, Rongchao Xu, Tian Tang, Guang Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SI

arXiv:2508.04780v2 Announce Type: replace Abstract: The increasing frequency of extreme weather events, such as hurricanes, highlights the urgent need for efficient and equitable power system restoration. Many electricity providers make restoration decisions primarily based on the volume of power re...

📖 Read original article


179. HCRide: Harmonizing Passenger Fairness and Driver Preference for Human-Centered Ride-Hailing ​

Author: Lin Jiang, Yu Yang, Guang Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.SI

arXiv:2508.04811v2 Announce Type: replace Abstract: Order dispatch systems play a vital role in ride-hailing services, which directly influence operator revenue, driver profit, and passenger experience. Most existing work focuses on improving system efficiency in terms of operator revenue, which may...

📖 Read original article


180. Communication-Enhanced Tutoring for Efficient Decentralized Multi-Agent Reinforcement Learning ​

Author: Maciej Wojtala, Bogusz Stefa'nczyk, Dominik Bogucki, {\L}ukasz Lepak, Pawe{\l} Wawrzy'nski
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.MA

arXiv:2508.13661v4 Announce Type: replace Abstract: Centralized Training with Decentralized Execution (CTDE) is the dominant paradigm in multi-agent reinforcement learning (MARL), enabling agents to act independently at test time while leveraging additional information during training. However, the ...

📖 Read original article


181. CountTRuCoLa: Rule Learning for Interpretable Temporal Knowledge Graph Forecasting ​

Author: Julia Gastinger, Christian Meilicke, Heiner Stuckenschmidt
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.09474v3 Announce Type: replace Abstract: We address the task of temporal knowledge graph forecasting with an inherently interpretable method based on symbolic rules. Motivated by recent work proposing a strong baseline based on recurrent facts, our approach learns four simple rule types, ...

📖 Read original article


182. Learning Neural Networks by Neuron Pursuit ​

Author: Akshay Kumar, Jarvis Haupt
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2509.12154v2 Announce Type: replace Abstract: The first part of this paper studies the evolution of gradient flow for homogeneous neural networks near a class of saddle points exhibiting a sparsity structure. The choice of these saddle points is motivated from previous works on homogeneous net...

📖 Read original article


183. Echo Flow Networks ​

Author: Hongbo Liu, Jia Xu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.24122v3 Announce Type: replace Abstract: At the heart of time-series forecasting (TSF) lies a fundamental challenge: how can models efficiently and effectively capture long-range temporal dependencies across ever-growing sequences? While deep learning has brought notable progress, convent...

📖 Read original article


184. Prototype-based Self-Supervised Multimodal Learning for PPG and Accelerometry Signals ​

Author: Wanting Mao, Maxwell A Xu, Harish Haresamudram, Mithun Saha, Santosh Kumar, James Matthew Rehg
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.09764v2 Announce Type: replace Abstract: Modeling multi-modal time-series data is critical for capturing system-level dynamics, particularly in biosignals where modalities such as ECG, PPG, EDA, and accelerometry provide complementary perspectives on interconnected physiological processes...

📖 Read original article


185. Contrastive Diffusion Alignment: Learning Structured Latents for Controllable Generation ​

Author: Ruchi Sandilya, Sumaira Perez, Charles Lynch, Lindsay Victoria, Benjamin Zebley, Derrick Matthew Buchanan, Mahendra T. Bhati, Nolan Williams, Timothy J. Spellman, Faith M. Gunning, Conor Liston, Logan Grosenick
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.14190v3 Announce Type: replace Abstract: Diffusion models excel at generation, but their latent spaces are high dimensional and not explicitly organized for interpretation or control. We introduce ConDA (Contrastive Diffusion Alignment), a plug-and-play geometry layer that applies contras...

📖 Read original article


186. Bi-Level Reinforcement Learning Pathway for Sim-to-Real Optimality ​

Author: Akhil S Anand, Shambhuraj Sawant, Paavo Parmas, Jasper Hoffmann, Dirk Reinhardt, Sebastien Gros
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.17709v2 Announce Type: replace Abstract: Training Reinforcement Learning (RL) policies using simulation models before deployment in real-world environments is a common strategy when real-world interaction is expensive. This approach is used in sim-to-real RL and in dyna-style model-based ...

📖 Read original article


187. Multicalibration Yields Better Matchings ​

Author: Riccardo Colini Baldeschi, Simone Di Gregorio, Simone Fioravanti, Federico Fusco, Ido Guy, Daniel Haimovich, Stefano Leonardi, Fridolin Linder, Lorenzo Perini, Matteo Russo, Cem Sirin, Niek Tax
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2511.11413v2 Announce Type: replace Abstract: Consider the problem of finding the best matching in a weighted graph where we only have access to predictions of the actual stochastic weights, based on an underlying context. If the predictor is the Bayes optimal one, then computing the best matc...

📖 Read original article


188. Stabilizing Multi-Attack Adversarial Training via Bandit Optimization ​

Author: Rui Wang, Zeming Wei, Xiyue Zhang, Meng Sun
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR, cs.CV, math.OC

arXiv:2511.12265v2 Announce Type: replace Abstract: Deep Neural Networks (DNNs) remain vulnerable to diverse adversarial perturbations, motivating multi-attack adversarial training (AT) for improved robustness. However, existing methods either incur prohibitive overhead by computing all attacks at e...

📖 Read original article


189. DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models ​

Author: Cheng Yin, Yankai Lin, Wang Xu, Sikyuen Tam, Xiangrui Zeng, Zhiyuan Liu, Zhouping Yin
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO

arXiv:2511.15669v3 Announce Type: replace Abstract: Does Chain-of-Thought (CoT) reasoning genuinely improve Vision Language Action (VLA) models, or does it merely add overhead? Existing CoT-VLA systems report limited and inconsistent gains, yet no prior work has rigorously diagnosed when and why CoT...

📖 Read original article


190. Interpreting GFlowNets for Drug Discovery: What probes can and cannot show ​

Author: Amirtha Varshini A S, Duminda S. Ranasinghe, Hok Hei Tam
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.BM

arXiv:2511.19264v2 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) construct molecules through sequential decisions, but their internal policies remain opaque, limiting adoption in drug discovery, where chemists need interpretable rationales for proposed structures. We present ...

📖 Read original article


191. A Mechanistic Analysis of Transformers for Dynamical Systems ​

Author: Gregory Duth'e, Nikolaos Evangelou, Wei Liu, Ioannis G. Kevrekidis, Eleni Chatzi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CE

arXiv:2512.21113v2 Announce Type: replace Abstract: Transformers are increasingly adopted for modeling and forecasting time-series, yet their internal mechanisms remain poorly understood from a dynamical systems perspective. In contrast to classical autoregressive and state-space models, which benef...

📖 Read original article


192. HERO: Hierarchical Evidential Reasoning Optimization for Radiology Report Generation via Reason-then-Summarize ​

Author: Kun Zhao, Guodong Liu, Hui Ji, Siyuan Dai, Pan Wang, Jifeng Song, Chenghua Lin, Liang Zhan, Haoteng Tang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.03321v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have substantially advanced Radiology Report Generation (RRG), yet aligning them through reinforcement learning (RL) remains challenging due to heterogeneous medical supervision. Vanilla Group Relative Polic...

📖 Read original article


193. RingSQL: Schema-Independent Synthetic Data Generation for Text-to-SQL Reinforcement Learning ​

Author: Marko Sterbentz, Kevin Cushing, Cameron Barrie, Kristian J. Hammond
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2601.05451v2 Announce Type: replace Abstract: Recent advances in text-to-SQL have been driven by larger models, better datasets, and new training methods like RLVR. However, progress remains limited by scarce high-quality training data, a problem RLVR is especially sensitive to since noisy dat...

📖 Read original article


194. Beyond the Dirac Delta: Mitigating Diversity Collapse in Reinforcement Fine-Tuning for Versatile Image Generation ​

Author: Jinmei Liu, Haoru Li, Zhenhong Sun, Chaofeng Chen, Yatao Bian, Bo Wang, Daoyi Dong, Chunlin Chen, Zhi Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.12401v2 Announce Type: replace Abstract: Reinforcement learning (RL) has emerged as a powerful paradigm for fine-tuning large-scale generative models, such as diffusion and flow models, to align with complex human preferences and user-specified tasks. A fundamental limitation remains \tex...

📖 Read original article


195. Distributional Active Inference ​

Author: Abdullah Akg"ul, Gulcin Baykal, Manuel Hau{\ss}mann, Mustafa Mert \c{C}elikok, Melih Kandemir
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.20985v2 Announce Type: replace Abstract: Optimal control of complex environments with robotic systems faces two complementary and intertwined challenges: efficient organization of sensory state information and far-sighted action planning. Because the reinforcement learning framework addre...

📖 Read original article


196. Efficient Training of Boltzmann Generators Using Off-Policy Log-Dispersion Regularization ​

Author: Henrik Schopmans, Christopher von Klitzing, Pascal Friederich
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.03729v3 Announce Type: replace Abstract: Sampling from unnormalized probability densities is a central challenge in computational science. Boltzmann generators are generative models that enable independent sampling from the Boltzmann distribution of physical systems at a given temperature...

📖 Read original article


197. Data-Aware and Scalable Sensitivity Analysis for Decision Tree Ensembles ​

Author: Namrita Varshney, Ashutosh Gupta, Arhaan Ahmad, Tanay V. Tayal, S. Akshay
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2602.07453v2 Announce Type: replace Abstract: Decision tree ensembles are widely used in critical domains, making robustness and sensitivity analysis essential to their trustworthiness. We study the feature sensitivity problem, which asks whether an ensemble is sensitive to a specified subset ...

📖 Read original article


198. RooflineBench: A Benchmarking Framework for On-Device LLMs via Roofline Analysis ​

Author: Zhen Bi, Xueshu Chen, Luoyang Sun, Yuhang Yao, Qing Shen, Jungang Lou, Cheng Deng
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR, cs.PF

arXiv:2602.11506v4 Announce Type: replace Abstract: The transition toward localized intelligence through Small Language Models (SLMs) has intensified the need for rigorous performance characterization on resource-constrained edge hardware. However, objectively measuring the theoretical performance c...

📖 Read original article


199. RiboSphere: Learning Unified and Efficient Representations of RNA Structures ​

Author: Zhou Zhang, Hanqun Cao, Cheng Tan, Fang Wu, Pheng Ann Heng, Tianfan Fu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.19636v2 Announce Type: replace Abstract: Accurate RNA structure modeling remains difficult because RNA backbones are highly flexible, non-canonical interactions are prevalent, and experimentally determined 3D structures are comparatively scarce. We introduce RiboSphere, a framework that l...

📖 Read original article


200. The Luna Bound Propagator for Formal Analysis of Neural Networks ​

Author: Henry LeCates, Haoze Wu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.LO

arXiv:2603.23878v3 Announce Type: replace Abstract: The parameterized CROWN analysis, a.k.a., alpha-CROWN has emerged as a practically successful abstract interpretation method for neural network verification. However, existing implementations of alpha-CROWN are limited to Python, which complicates ...

📖 Read original article


201. MOON3.0: Reasoning-aware Multimodal Representation Learning for E-commerce Product Understanding ​

Author: Junxian Wu, Chenghan Fu, Zhanheng Nie, Daoze Zhang, Bowen Wan, Wanxian Guan, Chuan Yu, Jian Xu, Bo Zheng
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.IR

arXiv:2604.00513v3 Announce Type: replace Abstract: With the rapid growth of e-commerce, exploring general representations rather than task-specific ones has attracted increasing attention. Although recent multimodal large language models (MLLMs) have driven significant progress in product understan...

📖 Read original article


202. Koopman-Based Nonlinear Identification and Model Predictive Control of a Turbofan Engine ​

Author: David Grasev
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2604.01730v2 Announce Type: replace Abstract: This paper investigates Koopman operator-based approaches for multivariable control of a two-spool turbofan engine. A physics-based component-level model is developed to generate training data and validate the controllers. A meta-heuristic extended...

📖 Read original article


203. Just Repair: A Minimal Denoising Network for Time Series Anomaly Detection ​

Author: Kadir-Kaan "Ozer, Ren'e Ebeling, Markus Enzweiler
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.17388v3 Announce Type: replace Abstract: Time series anomaly detectors have grown steadily more complex, incorporating attention mechanisms, adversarial training, and stochastic latent variables. Yet, it is unclear how much of this machinery detection actually requires. We test this quest...

📖 Read original article


204. Stable GFlowNets with TV Monitoring and Probabilistic Guarantees ​

Author: Zengxiang Lei, Ananth Shreekumar, Jonathan Rosenthal, Ruoyu Song, Alvaro A. Cardenas, Daniel J. Fremont, Dongyan Xu, Satish Ukkusuri, Z. Berkay Celik
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2605.01729v2 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) learn to sample states proportional to an unnormalized reward. Despite their theoretical promise, practical training is often unstable, exhibiting severe loss spikes and mode collapse. To tackle this, we first a...

📖 Read original article


205. Stable Attention Response for Reliable Precipitation Nowcasting ​

Author: Penghui Wen, Zexin Hu, Sen Zhang, Patrick Filippi, Xiaogang Zhu, Allen Benter, Thomas Bishop, Zhiyong Wang, Kun Hu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.13181v3 Announce Type: replace Abstract: Precipitation nowcasting remains challenging due to the highly localized, rapidly evolving, and heterogeneous nature of atmospheric dynamics. Although recent methods increasingly adopt attention-based architectures in both unimodal and multimodal s...

📖 Read original article


206. When Bits Break Recourse: Counterfactual-Faithful Quantization ​

Author: Chaymae Yahyati, Ismail Lamaakal, Khalid El Makkaoui, Ibrahim Ouahbi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2605.17160v4 Announce Type: replace Abstract: Model quantization is widely used to reduce memory, latency, and deployment cost, and is typically judged by whether predictive accuracy is preserved. In decision systems that provide algorithmic recourse, however, accuracy preservation is not suff...

📖 Read original article


207. Chebyshev Policies and the Mountain Car Problem: Reinforcement Learning for Low-Dimensional Control Tasks ​

Author: Stefan Huber, Hannes Unger, Georg Sch"afer, Jakob Rehrl
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.22305v5 Announce Type: replace Abstract: We analytically solve the Mountain Car problem, a canonical benchmark in RL, and derive an optimal control solution, closing a gap after 36 years. This enables us to reveal two surprising insights: The optimal control is quite simple, yet modern RL...

📖 Read original article


208. The Hamilton-Jacobi Theory of Deep Learning ​

Author: Jose Marie Antonio Mi~noza, Erika Fille T. Legara, Christopher P. Monterola
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DS, math.RT, physics.comp-ph

arXiv:2605.28983v2 Announce Type: replace Abstract: In this paper, training a neural network is identified, exactly, as a search through Hamilton--Jacobi initial-value problems: each gradient step selects the initial data of a viscous Hamilton--Jacobi equation whose Hopf--Cole propagator best fits t...

📖 Read original article


209. MemNovo: Look Back at the Spectrum for Balanced De Novo Peptide Sequencing from Mass Spectrometry ​

Author: Dongxin Lyu, Jingbo Zhou, Hongxin Xiang, Yuqiang Li, Jun Xia
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM

arXiv:2606.11868v2 Announce Type: replace Abstract: De novo peptide sequencing from tandem mass spectrometry is pivotal in proteomics, enabling identification of novel peptides without reference databases. While recent Transformer-based encoder-decoder models have achieved remarkable performance, we...

📖 Read original article


210. GRIMIP: A General Framework for Instance-Specific Configuration of MIP Solvers Using LLMs ​

Author: Yidong Luo, Xuemin Chen, Chenguang Wang, Fangzhou Zhu, Tao Zhong, Tianshu Yu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.23299v2 Announce Type: replace Abstract: Configuring the hyperparameters of Mixed-integer programming (MIP) solvers is a high-dimensional, instance-dependent optimization problem where suboptimal settings can degrade solving time by orders of magnitude. Default configurations are often su...

📖 Read original article


211. NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning ​

Author: Tianlin Pan, Lianyu Pang, Cheng Da, Huan Yang, Changqian Yu, Kun Gai, Wenhan Luo
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2606.27771v4 Announce Type: replace Abstract: Reinforcement learning (RL) post-training improves the reward alignment of flow-based generators, but often degrades perceptual quality in ways that are not captured by the reward proxy. We identify a simple structural signature of this drift: acro...

📖 Read original article


212. AdaBoosting Text Prompts for Vision-Language Models ​

Author: Seokhee Jin, Changhwan Sung, Sunung Mun, Hoyoung Kim, Jungseul Ok
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.00684v3 Announce Type: replace Abstract: The classification accuracy of pretrained Vision-Language Models (VLMs) relies on the quality of the text prompts. Handcrafted templates and Large Language Model (LLM)-generated descriptions not only make predictions more interpretable, but also en...

📖 Read original article


213. A More Accurate Algorithm Comparison through A/B Testing using Offline Evaluation Methods ​

Author: Koki Konishi, Masataka Ushiku, Yuta Saito
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.01958v2 Announce Type: replace Abstract: A/B testing is the gold standard for selecting the better algorithm in online services. While offline evaluation has attracted attention as a safer alternative due to the high experimental costs and the potential risk of degrading user experience a...

📖 Read original article


214. Foundations of Equivariant Deep Learning: Unifying Graph and Sheaf Neural Networks ​

Author: Yoshihiro Maruyama
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.03798v3 Announce Type: replace Abstract: Symmetry is everywhere in nature and society. Geometric deep learning builds architectures respecting group symmetries, whereas topological deep learning organizes computation through cells, incidence relations, and local-to-global structure. In th...

📖 Read original article


215. An interpretable Good--Turing restart criterion for k-means++ ​

Author: Renato Cordeiro de Amorim
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.08243v2 Announce Type: replace Abstract: The k-means++ algorithm is commonly restarted multiple times to avoid poor local optima, yet the number of restarts is almost always chosen arbitrarily and applied uniformly regardless of data set difficulty. This undermines any comparison relying ...

📖 Read original article


216. Forgetful Attention: An Auditable Support-Vector Memory for Selective Retention and Verified Deletion ​

Author: Vishwajith Ramesh
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12204v2 Announce Type: replace Abstract: Auditable memory requires a precise contract: which output is preserved, relative to which reference solve, and across which updates. We introduce Support Vector Attention (SV-Attention), a one-class support vector data description (SVDD) gate whos...

📖 Read original article


217. TerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale ​

Author: Zhouchonghao Wu, Akshay Rangesh, Weixin Li, Wei-Jer Chang, Zachary Lee, Saeed Bonab, Tim Wang, Wei Zhan
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO

arXiv:2607.13028v2 Announce Type: replace Abstract: Training robust autonomous driving agents requires a simulator fast enough for reinforcement learning at scale, realistic enough to ground behavior in real-world map structure, and diverse enough to cover the safety-critical long tail that logged d...

📖 Read original article


218. Optimizing the Preconditioner: A Black-box Online-to-Nonconvex Conversion with Static Regret Minimization Oracles ​

Author: Haichen Hu, David Simchi-Levi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2607.17607v2 Announce Type: replace Abstract: We study whether stochastic nonconvex optimization can be reduced to ordinary static regret minimization in online convex optimization in a black-box manner. For smooth nonconvex objectives, our reduction maintains a predictable gradient tracker, w...

📖 Read original article


219. Measuring the Dependency Gap: Diagnosing Inter-Column Fidelity in Tabular Generative Models ​

Author: Jie Zhang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.21636v3 Announce Type: replace Abstract: Synthetic tabular data are valued for preserving not just column-wise marginals but inter-column dependency, which carries much of the minority-class signal in domains such as fraud detection and clinical risk. Yet standard certification is largely...

📖 Read original article


220. Learning What Matters: Supervising Global Context Pruning with Causal Evidence Sets ​

Author: James E. Allchin
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.21692v3 Announce Type: replace Abstract: Pruning a long context means committing to the blocks a model will keep, and the usual selector is distilled from a dense teacher's attention. That assumes attention shows which context the answer depends on. We test the assumption on retrieval tas...

📖 Read original article


221. Wrong Design Intent Can Be Worse Than None: A Derangement-Control Diagnosis of Header Conditioning in CAD Program Completion ​

Author: Yang Xiao
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.23191v2 Announce Type: replace Abstract: Fine-tuned code LLMs are often conditioned on a design-intent header to steer parametric CAD generation, but whether the model reads that header's content has been tested neither under execution-level scoring nor with a causal control. We study CAD...

📖 Read original article


222. Breaking the Periodicity Assumption: Robust Tensorial Multi-View Clustering via Graph-Spectral Low-Rank Learning ​

Author: Jintian Ji, Xingsu Li, Songhe Feng
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25295v2 Announce Type: replace Abstract: Tensorial multi-view clustering (TMC) has achieved strong performance due to its ability to capture high-order correlations across multiple views. Most existing t-SVD-based TMC frameworks apply the Fast Fourier Transform (FFT) along the sample mode...

📖 Read original article


223. SE(3)-MeanFlow: Few-Step Protein Backbone Generation on Lie Groups ​

Author: Yikun Bai, Binghang Lu, Yikai Liu, Elaheh Akbari, Soheil Kolouri, Linxuan Wang, Ping He, Shuchan Wang, Ruqi Zhang, Guang Lin
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.27431v3 Announce Type: replace Abstract: Generative modeling of protein backbones promises the de novo design of proteins with prescribed structural and functional properties. Existing diffusion and flow-matching models produce high-quality backbones on SE(3)^N, but inference requires num...

📖 Read original article


224. Maglev: Sliding Recurrent Memory ​

Author: Bo Liu, Qiang Liu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02870v2 Announce Type: replace Abstract: We introduce \ours{}, a recurrent Transformer architecture with fixed-size memory that generalizes sliding-window attention while remaining parallelizable during training. \ours{} consists of two coupled models: a prefiller $Q$, which leverages ful...

📖 Read original article


225. Quantization Effects on Biomedical LLM Reliability ​

Author: Anton Rasmussen, Hong Qin
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03854v2 Announce Type: replace Abstract: When decoder language models are used as classifiers, predicted class probabilities depend on implementation choices, including the prompt template, verbalizer (label-to-token mapping), and scoring rule, that are rarely treated as experimental vari...

📖 Read original article


226. Latent Reward Registers for Diffusion Preference Alignment ​

Author: Yuanshen Guan, Zipeng Feng, Chengru Song, Zhiwei Xiong, Peiqin Sun
Published: 8/6/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.03929v2 Announce Type: replace Abstract: Aligning diffusion models with human preferences usually relies on a sparse terminal reward evaluated on the final generated samples, presenting a severe temporal credit-assignment challenge across the multi-step denoising process. We propose Laten...

📖 Read original article


227. E$^2$M: Double Bounded $\alpha$-Divergence Optimization for Tensor-based Discrete Density Estimation ​

Author: Kazu Ghalamkari, Jesper L{\o}ve Hinrich, Morten M{\o}rup
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2405.18220v4 Announce Type: replace-cross Abstract: Tensor-based discrete density estimation requires flexible modeling and proper divergence criteria to enable effective learning; however, traditional approaches using $\alpha$-divergence face analytical challenges due to the $\alpha$-power te...

📖 Read original article


228. Curiosity-Diffuser: Curiosity Guide Diffusion Models for Reliability ​

Author: Zihao Liu, Xing Liu, Yuhang Dong, Haitao Chang, Zhengxiong Liu, Panfeng Huang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2503.14833v2 Announce Type: replace-cross Abstract: One of the bottlenecks in robotic intelligence is the instability of neural network models. This leads to risks when applying intelligence in the physical world. Specifically, imitation policy based on neural network may generate hallucinatio...

📖 Read original article


229. Review Text as a Leading Indicator of Displayed Reputation in Platform Rating Systems: Evidence from 34 U.S. Short-Term Rental Markets ​

Author: Ali Safari
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CL, cs.LG

arXiv:2504.14053v2 Announce Type: replace-cross Abstract: Rating systems on accommodation platforms suffer from a familiar problem: nearly every listing displays a nearly perfect score, so the number that is supposed to separate good listings from bad ones barely varies. Whether the review text accu...

📖 Read original article


230. One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP ​

Author: Binyan Xu, Xilin Dai, Di Tang, Kehuan Zhang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2505.19840v3 Announce Type: replace-cross Abstract: Deep Neural Networks (DNNs) have achieved widespread success yet remain prone to adversarial attacks. Typically, such attacks either involve frequent queries to the target model or rely on surrogate models closely mirroring the target model -...

📖 Read original article


231. Esoteric Language Models: A Family of Any-Order Diffusion LLMs ​

Author: Subham Sekhar Sahoo, Zhihan Yang, Yash Akhauri, Johnna Liu, Deepansha Singh, Zhoujun Cheng, Zhengzhong Liu, Eric Xing, John Thickstun, Arash Vahdat
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2506.01928v5 Announce Type: replace-cross Abstract: Diffusion-based language models offer a compelling alternative to autoregressive (AR) models by enabling parallel and controllable generation. Within this family, Masked Diffusion Models (MDMs) currently perform best but still underperform AR...

📖 Read original article


232. MoCA: Multi-modal Cross-masked Autoencoder for Digital Health Measurements ​

Author: Howon Ryu, Yuliang Chen, Yacun Wang, Andrea Z. LaCroix, Chongzhi Di, Loki Natarajan, Yu Wang, Jingjing Zou
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP

arXiv:2506.02260v4 Announce Type: replace-cross Abstract: Wearable devices enable continuous multi-modal physiological and behavioral monitoring, yet analysis of these data streams faces fundamental challenges including the lack of gold-standard labels and incomplete sensor data. While self-supervis...

📖 Read original article


233. Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings for CLEF JOKER 2025 Task 2 ​

Author: Russell Taylor, Benjamin Herbert, Michael Sana
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.MA

arXiv:2507.06506v2 Announce Type: replace-cross Abstract: Translating wordplay across languages presents unique challenges that have long confounded both professional human translators and machine translation systems. This research proposes a novel approach for translating puns from English to Frenc...

📖 Read original article


234. Emergence of Hierarchical Emotion Organization in Large Language Models ​

Author: Maya Okawa, Bo Zhao, Eric J. Bigelow, Rose Yu, Tomer Ullman, Ekdeep Singh Lubana, Hidenori Tanaka
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2507.10599v3 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly power conversational agents, understanding how they model users' emotional states is critical for ethical deployment. Inspired by emotion wheels, i.e., a psychological framework that argues emotion...

📖 Read original article


235. The Yokai Learning Environment: Tracking Beliefs Over Space and Time ​

Author: Constantin Ruhdorfer, Matteo Bortoletto, Johannes Forkel, Jakob Foerster, Andreas Bulling
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2508.12480v3 Announce Type: replace-cross Abstract: The ability to cooperate with unknown partners is a central challenge in cooperative AI and widely studied in the form of zero-shot coordination (ZSC), which evaluates an algorithm by measuring the performance of independently trained agents ...

📖 Read original article


236. Integrated Noise and Safety Management in UAM via A Unified Reinforcement Learning Framework ​

Author: Surya Murthy, Zhenyu Gao, John-Paul Clarke, Ufuk Topcu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.MA, cs.LG

arXiv:2508.16440v2 Announce Type: replace-cross Abstract: Urban Air Mobility (UAM) envisions the widespread use of small aerial vehicles to transform transportation in dense urban environments. However, UAM faces critical operational challenges, particularly the balance between minimizing noise expo...

📖 Read original article


237. Arnold: A multi-task, multi-embodiment muscle transformer policy ​

Author: Boshi An, Alberto Silvio Chiappa, Merkourios Simos, Chengkun Li, Alexander Mathis
Published: 8/6/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, q-bio.QM

arXiv:2508.18066v2 Announce Type: replace-cross Abstract: Controlling high-dimensional and nonlinear musculoskeletal models of the human body is a foundational scientific challenge. Recent machine learning breakthroughs have heralded in-silico policies that master individual skills like reaching, ob...

📖 Read original article


238. Inferring Relative Consequences of Mechanical Ventilation from Observational Data Using Game-Based Comparisons ​

Author: David J. Albers, Tell D. Bennett, Jana de Wiljes, George Hripcsak, Bradford J. Smith, Peter D. Sottile, J. N. Stroh
Published: 8/6/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG, math.OC

arXiv:2510.15127v4 Announce Type: replace-cross Abstract: Identifying the effects of mechanical ventilation (MV) protocols in critical care requires analyzing data from heterogeneous patient-ventilator systems in the clinical decision-making environment. Multiscale interactions among these coupled c...

📖 Read original article


239. RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation ​

Author: Yuquan Xue, Guanxing Lu, Zhenyu Wu, Chuanrui Zhang, Bofang Jia, Zhengyi Gu, Ziwei Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2510.17640v4 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown strong manipulation capability when trained with large-scale imitation learning datasets. However, these datasets that predominantly consist of successful trajectories rarely provide the correcti...

📖 Read original article


240. Neural Diversity Regularizes Hallucinations in Language Models ​

Author: Kushal Chakrabarti, Nirmal Balachundhar
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2510.20690v3 Announce Type: replace-cross Abstract: Language models continue to hallucinate despite increases in parameters, compute, and data. We propose neural diversity -- decorrelated parallel representations -- as a principled mechanism that reduces hallucination rates at fixed parameter ...

📖 Read original article


241. Reinforcement Learning and Consumption-Savings Behavior ​

Author: Brandon Kaplowitz
Published: 8/6/2026, 4:00:00 AM
Categories: econ.GN, cs.AI, cs.LG, q-fin.EC

arXiv:2510.20748v2 Announce Type: replace-cross Abstract: This paper demonstrates how reinforcement learning can explain two puzzling empirical patterns in household consumption behavior during economic downturns. I develop a model where agents use Q-learning with neural network approximation to mak...

📖 Read original article


242. Model Inversion meets Cryptographic Fuzzy Extractors ​

Author: Mallika Prabhakar, Louise Xu, Prateek Saxena
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2510.25687v4 Announce Type: replace-cross Abstract: Model inversion attacks pose an open challenge to privacy-sensitive applications that use machine learning (ML) models. For example, face authentication systems use modern ML models to compute embedding vectors from face images of the enrolle...

📖 Read original article


243. Group-Equivariant Diffusion Models for Lattice Field Theory ​

Author: Octavio Vega, Javad Komijani, Aida El-Khadra, Marina Marinkovic
Published: 8/6/2026, 4:00:00 AM
Categories: hep-lat, cs.LG

arXiv:2510.26081v2 Announce Type: replace-cross Abstract: Near the critical point, Markov Chain Monte Carlo (MCMC) simulations of lattice quantum field theories (LQFT) become increasingly inefficient due to critical slowing down. In this work, we investigate score-based symmetry-preserving diffusion...

📖 Read original article


244. Convergence and Stability Analysis of Self-Consuming Generative Models with Heterogeneous Human Curation ​

Author: Hongru Zhao, Jinwen Fu, Tuan Pham
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2511.09002v3 Announce Type: replace-cross Abstract: Self-consuming generative models have received significant attention over the last few years. In this paper, we study a self-consuming generative model with heterogeneous preferences that is a generalization of the model in Ferbach et al. (20...

📖 Read original article


245. Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens ​

Author: Yiming Qin, Bomin Wei, Jiaxin Ge, Konstantinos Kallidromitis, Stephanie Fu, Trevor Darrell, XuDong Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2511.19418v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) excel at reasoning in linguistic space but struggle with perceptual understanding that requires dense visual perception, e.g., spatial reasoning and geometric awareness. This limitation stems from the fact that c...

📖 Read original article


246. MODEST: Multi-Optics Depth-of-Field Stereo Dataset ​

Author: Nisarg K. Trivedi, Vinayaka A. Belludi, Li-Yun Wang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV

arXiv:2511.20853v4 Announce Type: replace-cross Abstract: Training and evaluation of state-of-the-art computer vision algorithms for reliable shallow depth of field (DoF) rendering and defocus deblurring remain constrained by a persistent lack of large-scale, full-frame, high fidelity, real-image da...

📖 Read original article


247. BIM-Native Tokenization for Constraint-Aware Room Layout Synthesis ​

Author: Manuel Ladron de Guevara, Jinmo Rhee, Ardavan Bidgoli, Vaidas Razgaitis, Michael Bergin
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.LG

arXiv:2512.04832v3 Announce Type: replace-cross Abstract: We present a BIM-native tokenization for room-level layout synthesis in Building Information Modeling (BIM) scenes. The core contribution is representational: we encode each room as a sequence of BIM-Token Bundles, realized as columns of a sp...

📖 Read original article


248. Galaxy Phase-Space and Field-Level Cosmology: The Strength of Semi-Analytic Models ​

Author: Natal'i S. M. de Santi, Francisco Villaescusa-Navarro, Pablo Araya-Araya, Gabriella De Lucia, Fabio Fontanot, Lucia A. Perez, Manuel Arn'es-Curto, Violeta Gonzalez-Perez, 'Angel Chandro-G'omez, Rachel S. Somerville, Tiago Castro
Published: 8/6/2026, 4:00:00 AM
Categories: astro-ph.CO, astro-ph.GA, cs.LG

arXiv:2512.10222v2 Announce Type: replace-cross Abstract: Semi-analytic models are a widely used approach to simulate galaxy properties within a cosmological framework, relying on simplified yet physically motivated prescriptions. They have also proven to be an efficient alternative for generating a...

📖 Read original article


249. PSI3D: Plug-and-Play 3D Stochastic Inference with Slice-wise Latent Diffusion Prior ​

Author: Wenhan Guo, Jinglun Yu, Yaning Wang, Jin U. Kang, Yu Sun
Published: 8/6/2026, 4:00:00 AM
Categories: eess.IV, cs.LG

arXiv:2512.18367v2 Announce Type: replace-cross Abstract: Diffusion models are highly expressive image priors for Bayesian inverse problems. However, most diffusion models cannot operate on large-scale, high-dimensional data due to high training and inference costs. In this work, we introduce a Plug...

📖 Read original article


250. Fundamentals of quantum Boltzmann machine learning with visible and hidden units ​

Author: Mark M. Wilde
Published: 8/6/2026, 4:00:00 AM
Categories: quant-ph, cond-mat.stat-mech, cs.LG

arXiv:2512.19819v2 Announce Type: replace-cross Abstract: One of the primary applications of classical Boltzmann machines is generative modeling, wherein the goal is to tune the parameters of a model distribution so that it closely approximates a target distribution. Training relies on estimating th...

📖 Read original article


251. Energy-Tweedie: Score meets Score, Energy meets Energy ​

Author: Andrej Leban
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2512.23818v2 Announce Type: replace-cross Abstract: Denoising and score estimation are classically linked through Tweedie's formula, which relates the posterior mean under Gaussian noise to the Stein score of the noisy marginal. In this work, we extend this perspective beyond Gaussian noise to...

📖 Read original article


252. From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs ​

Author: Usha Shrestha, Dmitry Ignatov, Radu Timofte
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2601.03808v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved notable performance in code synthesis; however, data-aware augmentation remains a limiting factor, handled via heuristic design or brute-force approaches. We introduce a performance-aware, closed-loo...

📖 Read original article


253. Event Driven Clustering Algorithm ​

Author: David El-Chai Ben-Ezra, Adar Tal, Daniel Brisk
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2602.00115v3 Announce Type: replace-cross Abstract: This paper introduces a novel asynchronous, event-driven algorithm for real-time detection of small event clusters in event camera data. Similar to hierarchical agglomerative clustering methods, the proposed algorithm detects clusters based o...

📖 Read original article


254. Multi-Task GRPO: Reliable LLM Reasoning Across Tasks ​

Author: Shyam Sundhar Ramesh, Xiaotong Ji, Matthieu Zimmer, Sangwoong Yoon, Zhiyong Wang, Haitham Bou Ammar, Aurelien Lucchi, Ilija Bogunovic
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2602.05547v3 Announce Type: replace-cross Abstract: RL-based post-training with GRPO is widely used to improve large language models on individual reasoning tasks. However, real-world deployment requires reliable performance across diverse tasks. A straightforward multi-task adaptation of GRPO...

📖 Read original article


255. Non-Stationary Inventory Control with Lead Times ​

Author: Nele H. Amiri, Sean R. Sinclair, Maximiliano Udenio
Published: 8/6/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ML

arXiv:2602.05799v2 Announce Type: replace-cross Abstract: We study non-stationary single-item, periodic-review inventory control problems in which the demand distribution is unknown and may change over time. We analyze how demand non-stationarity affects learning performance across inventory models,...

📖 Read original article


256. Wedge Sampling: Efficient Tensor Completion with Nearly-Linear Sample Complexity ​

Author: Hengrui Luo, Anna Ma, Ludovic Stephan, Yizhe Zhu
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.NA, math.PR, math.ST, stat.TH

arXiv:2602.05869v3 Announce Type: replace-cross Abstract: We introduce Wedge Sampling, a new non-adaptive sampling scheme for low-rank tensor completion. We study recovery of an order-$k$ low-rank tensor of dimension $n\times\cdots\times n$ from structured observations of its entries. Unlike the sta...

📖 Read original article


257. Can Post-Training Transform LLMs into Causal Reasoners? ​

Author: Junqi Chen, Sirui Chen, Chaochao Lu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2602.06337v2 Announce Type: replace-cross Abstract: Causal inference is essential for decision-making but remains challenging for non-experts. While large language models (LLMs) show promise in this domain, their precise causal estimation capabilities are still limited, and the impact of post-...

📖 Read original article


258. MemFly: On-the-Fly Memory Optimization via Information Bottleneck ​

Author: Zhenyuan Zhang, Xianzhang Jia, Zhiqin Yang, Zhenbo Song, Wei Xue, Sirui Han, Yike Guo
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2602.07885v2 Announce Type: replace-cross Abstract: Long-term memory enables large language model agents to tackle complex tasks through historical interactions. However, existing frameworks encounter a fundamental dilemma between compressing redundant information efficiently and maintaining p...

📖 Read original article


259. stratum: A System Infrastructure for Massive Agent-Centric ML Workloads ​

Author: Arnab Phani, Elias Strauss, Sebastian Schelter
Published: 8/6/2026, 4:00:00 AM
Categories: cs.DB, cs.LG

arXiv:2603.03589v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) transform how machine learning (ML) pipelines are developed and evaluated. LLMs enable a new type of workload, agentic pipeline search, in which autonomous or semi-autonomous agents generate, va...

📖 Read original article


260. Simultaneous estimation of multiple discrete unimodal distributions under stochastic order constraints ​

Author: Yasuhiro Yoshida, Noriyoshi Sukegawa, Jiro Iwanaga
Published: 8/6/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ME

arXiv:2603.11532v2 Announce Type: replace-cross Abstract: We study the problem of estimating multiple discrete unimodal distributions, motivated by search behavior analysis on a real-world platform. To incorporate prior knowledge of precedence relations among distributions, we impose stochastic orde...

📖 Read original article


261. Imitation Learning from Human Motion Alone Does Not Guarantee Biomechanically Plausible Gait Kinetics ​

Author: Xinyi Liu, Jangwhan Ahn, Edgar Lobaton, Jennie Si, He Huang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2603.12408v3 Announce Type: replace-cross Abstract: Motion imitation learning (IL) is increasingly used in robotics and human gait modeling, yet its ability to recover biomechanically consistent joint moments without explicit kinetic information remains unclear. In this study, we examined whet...

📖 Read original article


262. Seeking Physics in Diffusion Noise ​

Author: Chujun Tang, Lei Zhong, Fangqiang Ding
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.RO

arXiv:2603.14294v3 Announce Type: replace-cross Abstract: Do video diffusion models encode signals predictive of physical plausibility? We probe intermediate denoising representations of pretrained Diffusion Transformers (DiTs) and find that physically plausible and implausible videos are partially ...

📖 Read original article


263. ReCodeAgent: A Multi-agent Workflow for Language-Agnostic Translation and Validation of Large-Scale Repositories ​

Author: Ali Reza Ibrahimzada, Brandon Paulsen, Daniel Kroening, Reyhaneh Jabbarvand
Published: 8/6/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2604.07341v2 Announce Type: replace-cross Abstract: Most repository-level code translation and validation techniques have been evaluated on a single source-target programming language (PL) pair, owing to the complex engineering effort required to adapt new PL pairs. Programming agents can enab...

📖 Read original article


264. SHIELD: A Segmented Hierarchical Memory Architecture for Energy-Efficient LLM Inference on Edge NPUs ​

Author: Jintao Zhang, Xuanyao Fong
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AR, cs.LG

arXiv:2604.07396v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) inference on edge Neural Processing Units (NPUs) is fundamentally constrained by limited on-chip memory capacity. Although high-density embedded DRAM (eDRAM) is attractive for storing activation workspaces, its peri...

📖 Read original article


265. From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs ​

Author: Itay Itzhak, Eliya Habba, Gabriel Stanovsky, Yonatan Belinkov
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.14137v3 Announce Type: replace-cross Abstract: Evaluating LLMs is challenging, as benchmark scores often fail to capture models' real-world usefulness. Instead, users often rely on ``vibe-testing'': informal experience-based evaluation, such as comparing models on coding tasks related to ...

📖 Read original article


266. Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models ​

Author: Danae S'anchez Villegas, Samuel Lewis-Lim, Nikolaos Aletras, Desmond Elliott
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CV, cs.LG

arXiv:2604.14888v3 Announce Type: replace-cross Abstract: Recent advances in vision language models (VLMs) offer reasoning capabilities, yet how these unfold and integrate visual and textual information remains unclear. We analyze reasoning dynamics in 18 VLMs covering instruction-tuned and reasonin...

📖 Read original article


267. DeepImagine: Clinical Trial Outcome Prediction via Stepwise Local Counterfactual Imaginations ​

Author: Youze Zheng, Jianyou Wang, Yuhan Chen, Matthew Feng, Longtian Bao, Hanyuan Zhang, Maxim Khan, Aditya K. Sehgal, Christopher D. Rosin, Umber Dube, Ramamohan Paturi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.23054v2 Announce Type: replace-cross Abstract: Predicting the outcomes of prospective clinical trials remains a major challenge. Clinical trial outcomes result from complex interactions among experimental factors such as drug interventions, participant demographics, and protocols. Here, w...

📖 Read original article


268. Leakage-Audited Benchmarking Reveals Limited Evidence for Cross-Subject Auditory-Evoked EEG Vowel Perception Decoding ​

Author: Xiaoyang Li, Zeyan Tao
Published: 8/6/2026, 4:00:00 AM
Categories: eess.SP, cs.CL, cs.CV, cs.LG, cs.SD, q-bio.NC

arXiv:2605.00865v2 Announce Type: replace-cross Abstract: We tested whether auditory-evoked EEG supports subject-independent five-vowel perception decoding when trial identity, model identity, prediction provenance, and participant-level inference are controlled within a single benchmark. We reconst...

📖 Read original article


269. A Comprehensive Evaluation of Code Language Models for Security Patch Detection ​

Author: Nils Loose, Joseph Bienh"uls, Kristoffer Hempel, Felix M"achtle, Thomas Eisenbarth
Published: 8/6/2026, 4:00:00 AM
Categories: cs.SE, cs.CR, cs.LG

arXiv:2605.13138v2 Announce Type: replace-cross Abstract: Automated detection of vulnerability-fixing commits (\vfcs) is critical for timely security patch deployment, as advisory databases lag patch releases by a median of 25 days and many fixes never receive advisories. Code language models are in...

📖 Read original article


270. Spectral Integrated Gradients for Coarse-to-Fine Feature Attribution ​

Author: Soyeon Kim, Seongwoo Lim, Kyowoon Lee, Jaesik Choi
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2605.19607v2 Announce Type: replace-cross Abstract: Integrated Gradients (IG) is a widely adopted feature attribution method that satisfies desirable axiomatic properties. However, the choice of integration path significantly affects the quality of attributions, and the standard straight-line ...

📖 Read original article


271. Distributionally Robust Transfer Learning with Structurally Missing Covariates, with Application to Cross-National Cardiac Arrest Prediction ​

Author: Siqi Li, Chuan Hong, Ziye Tian, Benjamin Sieu-Hon Leong, Koshi Nakagawa, Hideharu Tanaka, Sang Do Shin, Khuong Quoc Dai, Do Ngoc Son, Marcus Eng Hock Ong, Nan Liu, Molei Liu
Published: 8/6/2026, 4:00:00 AM
Categories: stat.AP, cs.AI, cs.LG, stat.ML

arXiv:2605.24212v2 Announce Type: replace-cross Abstract: Deploying clinical prediction models across healthcare systems often fails when key training covariates are unavailable at deployment and labeled outcomes are limited in the target domain. For example, high-performing models for out-of-hospit...

📖 Read original article


272. Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns ​

Author: Guni Sharon
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2605.28566v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, yet their standard generation process -- auto-regressive token prediction -- is inherently myopic and prone to cascading errors. To address this, the Tree-of-Th...

📖 Read original article


273. Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers ​

Author: Edward Y. Chang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.04421v3 Announce Type: replace-cross Abstract: Many agentic systems and LLM pipelines correct mistakes by optimizing outcome reward. This addresses only the what of failure; the why and when may go unlogged, allowing the same error to recur across episodes. We propose long-horizon tempora...

📖 Read original article


274. GENEB: Why Genomic Models Are Hard to Compare ​

Author: Daria Ledneva, Mikhail Nuridinov, Denis Kuznetsov
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, q-bio.GN

arXiv:2606.04525v4 Announce Type: replace-cross Abstract: Progress in genomic foundation models is difficult to assess due to fragmented benchmarks, incompatible evaluation protocols, and task-specific reporting. As a result, claims of superiority or generality across models are often not directly c...

📖 Read original article


275. Sequential Kernel-based Conditional Independence Testing via Adaptive Betting ​

Author: Zheng He, Danica J. Sutherland
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2606.18993v2 Announce Type: replace-cross Abstract: Testing conditional independence is fundamental yet intrinsically difficult: without additional assumptions, Type I error control is impossible in general. The "Model-X'' paradigm addresses this difficulty by assuming exact knowledge of a rel...

📖 Read original article


276. Best-of-$N$ TTS Evaluation is Confounded by ASR Family Alignment ​

Author: Taehyung Yu, Seongjae Kang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.SD

arXiv:2607.08256v2 Announce Type: replace-cross Abstract: Best-of-$N$ (BoN) inference improves content consistency in zero-shot text-to-speech by selecting among multiple candidates with an automatic speech recognition (ASR) verifier. We identify an evaluation confound: the apparent quality of a ver...

📖 Read original article


277. Leveraging Interpretable Tsetlin Machine for PDF Malware Detection ​

Author: Rahul Jaiswal, Ole-Christoffer Granmo
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2607.09290v2 Announce Type: replace-cross Abstract: In the digital era, Portable Document Format (PDF) is one of the most widely used file formats for storing and exchanging digital documents due to its platform independence and rich functionality. However, these same capabilities have also ma...

📖 Read original article


278. Amplitude-Only FFN Intervention for Tool-Structured LLM Inference Method: Gated Evaluation Protocol, and Cross-Model Empirical Results ​

Author: Sheng Xu, Zhen Chen, Junhua Wang, Boyuan Huang, Ke Jia, Jiadun Zhu, Yiming Xu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.11183v3 Announce Type: replace-cross Abstract: Large language models increasingly operate as tool-using agents, where small format, argument, or function-call errors can invalidate otherwise plausible responses. We study inference-time feed-forward network (FFN) intervention as a way to i...

📖 Read original article


279. Theory-Level Autoformalization: From Isolated Statements to Unified Formal Knowledge Bases ​

Author: Marcus J. Min, Mike He, Zhaoyu Li, Zixuan Yi, Sharad Malik, Aarti Gupta, Xujie Si, Osbert Bastani
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.PL

arXiv:2607.13292v2 Announce Type: replace-cross Abstract: Autoformalization translates informal natural language into formal, machine-verifiable languages. While most work focuses on individual statements, real formalization efforts are inherently theory-level: they require an entire web of axioms, ...

📖 Read original article


280. Decision Making Needs Uncertainty Quantification [Lecture Notes] ​

Author: Osvaldo Simeone
Published: 8/6/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, math.IT

arXiv:2607.14407v3 Announce Type: replace-cross Abstract: Many signal processing systems ultimately exist to {act}. Whenever the state variable that determines the action to be taken by a decision maker, or agent, is uncertain, the way that uncertainty is represented decides how well the agent perfo...

📖 Read original article


281. Mixing-Free and Signal-Optimal Learning of Gaussian Graphical Models from Glauber Dynamics ​

Author: Vignesh Tirukkonda, Gautam Dasarathy
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2607.18559v2 Announce Type: replace-cross Abstract: Gaussian graphical model selection is usually studied under independent sampling, but in many applications the data arise as a single trajectory of a dependent stochastic process. We study exact recovery of the graph from one trajectory of ra...

📖 Read original article


282. Plausibility-Driven Prioritization of Candidate Biomedical Annotations ​

Author: Emanuele Cavalleri, Miad Alavinezhad, Dario Malchiodi, Marco Mesiti
Published: 8/6/2026, 4:00:00 AM
Categories: q-bio.QM, cs.DB, cs.LG

arXiv:2607.20163v2 Announce Type: replace-cross Abstract: The rapid growth of biomedical knowledge has made the validation of automatically generated biological annotations a major bottleneck in biomedical curation. While computational methods can rapidly produce large numbers of candidate annotatio...

📖 Read original article


283. Guarantees by Construction for Learned Finite Volume Schemes on Steady Supersonic Flow ​

Author: Denis Gueyffier (ONERA -- Institut Polytechnique de Paris)
Published: 8/6/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG, cs.NA, math.NA

arXiv:2607.20171v3 Announce Type: replace-cross Abstract: A second order finite volume scheme rests on two local quantities: a gradient reconstructed in each cell, and a limiter which scales it down where the reconstruction would overshoot. Both are set by fixed formulas, and on coarse unstructured ...

📖 Read original article


284. Cautious optimism for deep parameterized quantum circuits ​

Author: Marie Kempkes, Elies Gil-Fuster, Carlos Bravo-Prieto, Aroosa Ijaz, Alissa Wilms, Jens Eisert, Evert van Nieuwenburg, Vedran Dunjko
Published: 8/6/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, stat.ML

arXiv:2607.21409v2 Announce Type: replace-cross Abstract: A central challenge in quantum machine learning is understanding the scaling behavior of parameterized quantum circuits (PQCs). In particular, it remains unclear how their performance on unseen data changes as the number of trainable paramete...

📖 Read original article


285. Subject-Level Heterogeneity in EEG Motor Imagery Decoding: A Large-Scale Benchmark and Portfolio-Based Reduction of the Search Space ​

Author: Xavier Vasques, Paul Barbaste, Olivier Oullier
Published: 8/6/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG

arXiv:2607.22778v2 Announce Type: replace-cross Abstract: Robust EEG motor imagery decoding remains limited by strong inter-individual variability, making it difficult to identify pipelines that generalize across users. We present a large-scale, standardized within-session benchmark of decoding pipe...

📖 Read original article


286. SpecBox: Speculative Sandbox Scheduling for Efficient LLM Agent Serving ​

Author: Yihui Zhang (Beihang University), Tianyu Wo (Beihang University), Jinghao Wang (Beihang University), Xiaoyang Sun (University of Leeds), Menghao Zhang (Beihang University), Cangzhou Yuan (Beihang University), Li Li (Beihang University), Chunming Hu (Beihang University), Albert Y. Zomaya (The University of Sydney), Renyu Yang (Beihang University)
Published: 8/6/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG, cs.PF

arXiv:2607.23933v2 Announce Type: replace-cross Abstract: As LLM agents increasingly rely on the Model Context Protocol (MCP) to invoke isolated external sandboxes, disaggregated sandbox deployment introduces a fundamental tension between resource utilization and interactive tail latency. Persistent...

📖 Read original article


287. Look Ahead Before You Distill: Future Trajectory Validation of Teacher Guidance for Agentic On-Policy Distillation ​

Author: Chishui Chen, Yaoyou Fan, Te Sun, Yi Yang, Chenghao Sun, Delin Mao, Hongbo Qiao, Zuowei Zhang, Junxi Wang, Chenxing Sun, Yangen Hu, Lu Pan, Xuyang Liu, Linfeng Zhang
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.01953v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) provides teacher supervision on states visited by the student, reducing the distribution gap between training and inference. However, in multi-turn agentic tasks, student deviations may accumulate over time, gradu...

📖 Read original article


288. STEAM: A Spatio-TEmporal Alignment Mixture-of-Experts Model with Hierarchical Pre-training for EEG Decoding ​

Author: Zhu Chen, Dingkun Liu, Yuheng Chen, Dongrui Wu
Published: 8/6/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.02070v2 Announce Type: replace-cross Abstract: Brain-computer interfaces (BCIs) have been widely used in motor rehabilitation, disease diagnosis, and other neural engineering scenarios. However, conventional neural signal decoding algorithms often suffer from limited generalizability and ...

📖 Read original article


289. Instruction-Conditioned Exploration for Reinforcement Learning with Self-Distillation to an Unconditioned Policy ​

Author: Jim Dilkes, Vahid Yazdanpanah, Sebastian Stein
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.02087v2 Announce Type: replace-cross Abstract: Post-training Large Language Models (LLMs) with Reinforcement Learning (RL) has become an important tool for improving model capabilities, but the LLM action-space structure introduces challenges distinct from classical RL, with implications ...

📖 Read original article


290. When Correct Solutions Repeat: Rarity-Aware Credit Redistribution for GRPO ​

Author: Zhe Cao, Miaowen Wen, Fangjiong Chen
Published: 8/6/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.03467v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) com- monly optimizes each correct completion as an independent learning signal. In GRPO, this completion-level uniformity creates structure-level skew: recurring correct solution forms acc...

📖 Read original article


291. Robust Low-Tubal-Rank Tensor Completion under Cross-Concentrated Sampling ​

Author: HanQin Cai, Longxiu Huang, Jing Qin, Chengyue Wu
Published: 8/6/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, cs.NA, math.IT, math.NA

arXiv:2608.03928v2 Announce Type: replace-cross Abstract: Tensor cross-concentrated sampling (t-CCS) bridges entrywise sampling and t-CUR slice-wise sampling by observing entries only within selected horizontal and lateral slices. Existing t-CCS completion methods, however, assume that the observati...

📖 Read original article