Skip to content

arXiv cs.LG - 2026-08-27 ​

308 items collected.


1. Dynamic Influence-Weighted Distillation for Single-IMU Activity Recognition ​

Author: Bingxuan Xie
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.HC

arXiv:2608.24904v1 Announce Type: new Abstract: Inertial sensors at multiple body locations can improve activity recognition, but requiring every sensor at inference increases the deployment burden. We study whether four synchronized IMUs available during training can improve a student that uses onl...

📖 Read original article


Author: Surya Saka
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.24936v1 Announce Type: new Abstract: We present GreenLeaf Law Embed Tiny, a 0.6B parameter embedding model for legal domain retrieval. GreenLeaf-Tiny achieves 75.11% on the Massive Legal Embedding Benchmark (MLEB) and 64.38% on MTEB(Law, v1),demonstrating competitive performance among mod...

📖 Read original article


3. Multi-Modal Anomaly Detection: A Survey ​

Author: Xudong Mou, Zexin Wu, Chuan Luo, Shiru Chen, Xudong Liu, Chunming Hu, Renyu Yang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.24937v1 Announce Type: new Abstract: Multi-Modal Anomaly Detection (MMAD) detects rare abnormal events from heterogeneous data sources and is increasingly used in safety- and reliability-critical applications such as industrial inspection and cybersecurity. Yet the literature is fragmente...

📖 Read original article


4. ExFold: Unified Expert Folding for Training-Free MoE Prefill-Decode Acceleration ​

Author: Juntong Wu, Yifei Liu, Junyi Chen, Siqi Fan, Chaoran Feng, Minghao Li, Liujie Zhang, Weihang Chen, Li Yuan
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.24938v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models scale capacity for strong quality while keeping per-token compute bounded through sparse expert activation. Yet low-latency MoE serving is increasingly challenging, because it spans two inference phases with fundamentall...

📖 Read original article


5. When Does Frequency Decomposition Benefit Physics-Informed Neural Networks? A Preliminary Ablation Study ​

Author: Shubham Rai
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.24940v1 Announce Type: new Abstract: Partial differential equations (PDEs) often have high-frequency and multi-scale features that neural networks struggle to approximate. Physics-Informed Neural Networks (PINNs) build the governing equations directly into training, but suffer from spectr...

📖 Read original article


6. FAMPWQ: Fisher Information-based Adaptive Mixed Precision Weight Quantization for Effective LLM Inference ​

Author: Gongwei Lee, Ji Liu, Juncheng Jia, Ji Wu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC

arXiv:2608.24945v1 Announce Type: new Abstract: Recent years have witnessed remarkable achievements of Large Language Models (LLMs) in multiple domains, while the excessive resource requirements of LLMs hinder the deployment on resource-constrained devices. Although model quantization stands out as ...

📖 Read original article


7. MacroAgent: Regularity-Aware Macro Legalization with LLM-Agent-Designed Contour Algorithms ​

Author: Jiaxi Jiang, Xufeng Yao, Yuxuan Zhao, Yuntao Lu, Peiyu Liao, Zuodong Zhang, Yibo Lin, Bei Yu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.24946v1 Announce Type: new Abstract: Macros constitute a large part of the core area in modern very large-scale integration (VLSI) designs. Moreover, macro positions have a significant impact on the final quality of result (QoR), and macro legalization is typically the final step in deter...

📖 Read original article


8. CAT-GS: Balanced Multimodal Learning via Calibrated Gating and Fusion Surgery ​

Author: Mahir Shahriar Tamim, Sharjil Khan, Md. Samiul Alim, Tanvir Ahmed Khan, Shafin Rahman, Nabeel Mohammed
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.24947v1 Announce Type: new Abstract: End-to-end training of multimodal neural networks often exhibits unstable neural dynamics characterized by three coupled failure modes that degrade learning: (i) modality imbalance, where one branch dominates gradient-based optimization; (ii) unstable ...

📖 Read original article


9. Demystifying Reinforcement Learning Post-Training of Language Models ​

Author: Donovan Clay, Saket Gollapudi, Sankar Harilal, Min Jang, Jacob Morrison, Sewoong Oh, Natasha Jaques
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.24949v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has emerged as a powerful framework for enhancing the capabilities of large language models (LLMs), enabling impressive reasoning, math, and coding capabilities. Yet for many researchers and practitioners, the ...

📖 Read original article


10. AFDBench: A Reasoning-First AI Scientist for NationalWeather Service Forecast Discussions ​

Author: Manmeet Singh, Somnath Luitel, Prabhjot Singh, Manraaj Banga, Naveen Sudharsan, Josh Durkee
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2608.24954v1 Announce Type: new Abstract: Large language models (LLMs) hallucinate numerical values when generating high-stakes meteorological text, posing risks for weather communication. We present AFDBench, an AI meteorologist that generates professional Area Forecast Discussions (AFDs) by ...

📖 Read original article


11. Why and When Neural Networks Improve Local Approximation in Optimization ​

Author: Chengkuo Bian, Pengcheng Xie
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.24963v1 Announce Type: new Abstract: Published experience with neural surrogates in derivative-free optimisation is contradictory: the same family of models that cuts the evaluation count of one solver leaves another unchanged, or makes it worse. We show that the contradiction dissolves o...

📖 Read original article


12. Physics-Informed Error Field Learning: A Post-Training Optimization Framework for Physics-Informed Neural Networks ​

Author: Jiuyun Sun, Yong Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.24970v1 Announce Type: new Abstract: Physics-Informed Neural Networks (PINNs) have emerged as an important class of numerical methods for solving partial differential equations (PDEs). However, during the late-stage optimization process, further parameter updates often yield diminishing a...

📖 Read original article


13. Resource-Efficient Pruning for Transformer via Low-Rank Importance Estimation ​

Author: Peng Liu, Huibing Zeng, Yiqun Zhang, Yang Yi, Jigang Wu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET

arXiv:2608.24973v1 Announce Type: new Abstract: With the rapid development of large-scale pre-trained language models based on Transformer architectures, their high computational and memory costs have become a major obstacle to deployment, especially in resource-constrained environments. Traditional...

📖 Read original article


14. Clearing the Underbrush: AI-Enhanced RF Interference Suppression ​

Author: Rahul Jain, Pierre Trepagnier, Rick Gentile, Joey Botero, Alexia Schulz
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP

arXiv:2608.24974v1 Announce Type: new Abstract: AI-based structured interference rejection has grown more popular because deep learning approaches can outperform traditional methods by jointly considering the signal of interest (SOI) and the signal mixture (SOI plus interference). This work builds o...

📖 Read original article


15. MSR-IVA: Masked Structural Residual Independent Vector Analysis for State-Aware Fusion of Structural MRI and Dynamic Functional Network Connectivity ​

Author: Victor Solomon, Zening Fu, Rafal Angryk, Vince D. Calhoun, Jingyu Liu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.24978v1 Announce Type: new Abstract: Multimodal fusion of structural MRI (sMRI) and dynamic functional network connectivity (dFNC) can reveal how brain structure relates to changing functional states. When the same structural latent representation is coupled with multiple states, applying...

📖 Read original article


16. D$^3$-MOPD: Adaptive Dynamic Domain ScheDuling for Efficient Multi-Teacher Distillation ​

Author: Zechen Sun, Zhiwei Zhang, Fei Zhao, Juntao Li, Mu Chuan, Huayu Deng, Guojian Zhan, Wenliang Chen, Yao Hu, Min Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.24987v1 Announce Type: new Abstract: Multi-teacher on-policy distillation (MOPD) distills several domain-expert teachers into a single student by minimizing per-domain reverse-KL divergence on the student's own rollouts. Existing approaches typically fix the per-domain data mixture before...

📖 Read original article


17. Rollout-Decoded Reconstruction for Long-Horizon Prediction in Latent World Models ​

Author: Rishi Shah, Rishav Shrestha
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25017v1 Announce Type: new Abstract: A latent world model trains its decoder on latents anchored to observations, then deploys it on the model's own free-running rollout, hundreds of steps past the last observation. Rollout-Decoded Reconstruction (RDR) closes this gap with a single loss t...

📖 Read original article


18. On the Representational Geometry of Dynamic Programs ​

Author: Richard F. M. Lim, Ruriko Yoshida
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.DM, math.AG, math.CO

arXiv:2608.25034v1 Announce Type: new Abstract: Standard neural architectures often fail to generalize to longer inputs for dynamic programming (DP) targets. We investigate what makes this hard geometrically. Every finite min-plus DP is a shortest path on a DAG, which is equivalently a tropical poly...

📖 Read original article


19. DeMMO: Longitudinal and Cross-Disease Modelling of Digital Mobility Outcomes via Multi-Task Learning ​

Author: Menghui Zhou, Zhipeng Yuan, Vitaveska Lanfranchi, Po Yang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25073v1 Announce Type: new Abstract: Digital mobility outcomes (DMOs) derived from wearable sensors characterise mobility in daily life and offer a promising means of monitoring disease progression. Yet most DMO studies examine one disease at one visit; they do not model how multivariate ...

📖 Read original article


20. NVExplain: Explaining Time Series Forecasting with Latent Trajectory Analysis and Structure-Preserving Surrogates ​

Author: Muyan Anna Li, Manikandan Ravikiran, Aditi Gautam
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.25080v1 Announce Type: new Abstract: Time series forecasting models are widely used in high-stakes settings, yet their predictions remain difficult to interpret because existing post-hoc methods often ignore temporal dependence and fail to provide horizon-specific explanations. We propose...

📖 Read original article


21. The Frame Kernel Method for Multiscale Operator Learning ​

Author: Branden Frieden, Ryan Whitehead, M. Keith Ballard, Robert M. Kirby, Varun Shankar
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2608.25084v1 Announce Type: new Abstract: We present a natively multiscale operator learning method for the surrogate modeling of (numerical solvers for) multiscale partial differential equations (PDEs). The primary novelty of our method lies in a novel multiscale kernel frame function approxi...

📖 Read original article


22. The Von-Neumann State-Space Transformer for neural decoding ​

Author: Morteza Sarafyazd
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.NC

arXiv:2608.25088v1 Announce Type: new Abstract: Cortical computation is strikingly low-dimensional: a handful of latent variables, carried in a neural population's activity, steer the higher-dimensional responses of individual neurons. Our aim is sample efficiency-models that decode well from limite...

📖 Read original article


23. Understanding the Energy Scaling of Large Language Model Inference Across Context Lengths and Attention Architectures ​

Author: Molka Chkir, Syed Muhammad Danish, Jos H"oll, Arghavan Asad
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25096v1 Announce Type: new Abstract: The growing adoption of large language models (LLMs) has raised increasing concerns about the energy consumption and environmental impact of inference. This paper presents a systematic empirical study of decode-phase energy consumption across represent...

📖 Read original article


24. Flower Hub: A Reproducible Benchmarking Platform for Federated Learning in Simulation and Deployment ​

Author: Yan Gao, Mohammad Naseri, Javier Fernandez-Marques, Dimitris Stripelis, Lorenzo Sani, Davide Eynard, Fan Zhang, Hong Jia, Ting Dang, D. B. Emerson, Fatemeh Tavakoli, Ole Werger, Lars Wulfert, Petros Demetrakopoulos, Sofia Tsekeridou, InSeo Song, KangYoon Lee, Honghao Li, Lingjuan Lyu, John P Dickerson, Daniel Janes Beutel, Nicholas D. Lane
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25114v1 Announce Type: new Abstract: Federated learning (FL) has emerged as a key approach for training models across decentralized data, yet benchmarking in FL remains difficult to reproduce, compare, and extend. Existing evaluations are often tied to custom infrastructure, released as i...

📖 Read original article


25. GRAPE: Gradient Refinement and Progress-Aware Exploitation for Query-Efficient High-Dimensional Bayesian Optimization ​

Author: Richard Cornelius Suwandi, Feng Yin
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.25116v1 Announce Type: new Abstract: Optimizing expensive, high-dimensional black-box functions remains a central challenge in modern machine learning and scientific discovery. While local Bayesian optimization mitigates the curse of dimensionality, existing techniques often prioritize th...

📖 Read original article


26. Toward Machine Learning with the Unit as a Primitive: Learning from Unit-Linked Events ​

Author: Heyang Gong
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.25118v1 Announce Type: new Abstract: Machine learning is usually formalized through samples, while the persistent individual to which multiple observed or possible events refer often remains implicit. We propose the \emph{unit} as an explicit primitive at the level of task semantics. A le...

📖 Read original article


27. Multimodal Injury Risk Prediction in Tennis ​

Author: Francisco Erramuspe Alvarez, Shobharani Polasa, Weihao Qu, Jay Wang, Ling Zheng
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25126v1 Announce Type: new Abstract: Machine learning has had a significant positive impact on the prediction of athlete performance and injury risk. Most works in this field rely on subjective observations and expert assessments, which restrict their effectiveness. In sports like soccer,...

📖 Read original article


28. When Does Context Routing Help? A Systematic Study of Multi-Modal Fusion in Time Series Forecasting ​

Author: Ruizhe Zhou, Gaoyuan Du, Xiaoyang Liu, Haoqi Yao, Deepayan Chakrabarti, Jiating Lin, Yixuan Shen
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.AP

arXiv:2608.25128v1 Announce Type: new Abstract: Multi-modal time series forecasting methods integrate auxiliary context into temporal predictions through increasingly sophisticated fusion mechanisms. A growing body of work reports substantial gains, yet it is often unclear whether they reflect genui...

📖 Read original article


29. Rethinking the Transferable Adversarial Attacks and Robust Defense in Federated Learning ​

Author: Zuobin Xiong, Deval Mukherjee, Homook Cho, Wei Li
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.25133v1 Announce Type: new Abstract: The development of federated learning (FL) techniques has helped improve the privacy preservation of users' data and extended the applications of machine learning models. However, the involvement of a large number of users in FL also creates open oppor...

📖 Read original article


30. Drift Variation Autoencoder: Unifying Generation and Representation Learning through Conditional Posterior Flow Matching ​

Author: Jiarui Cao
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25138v1 Announce Type: new Abstract: Stochastic masking, cropping, or modality removal makes deterministic reconstruction an incomplete target: one observation can admit many clean completions. This work takes the corresponding posterior $P(X\mid C)$ as the common statistical object for c...

📖 Read original article


31. SNAP-KG: Streaming Node Assignment via Projection for Knowledge Graph Entity Integration ​

Author: Jui-Chien Lin, Mohammad Mohammadi Amiri, Oshani Seneviratne
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25149v1 Announce Type: new Abstract: Knowledge graph (KG) construction pipelines must continuously integrate newly arriving entities into a growing graph. Unlike inserting triples between existing nodes, a newly arriving entity has no graph connectivity: it emerges from the acquisition ph...

📖 Read original article


32. Bayesian Flow Networks for Offline Trajectory Planning ​

Author: Ludvig Killingberg, Helge Langseth
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25163v1 Announce Type: new Abstract: Offline reinforcement learning (RL) leverages static datasets to learn decision policies without real-time environment interaction. While recent sequence-modeling approaches rely on continuous diffusion models for trajectory synthesis, applying these m...

📖 Read original article


33. Simultaneous inference of environmental and interaction forces in collective dynamics ​

Author: Nipuni de Silva, Ming Zhong, James M. Greene
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, math.DS

arXiv:2608.25181v1 Announce Type: new Abstract: Collective dynamics arise in a wide range of physical, biological, and engineering applications. Examples include cell migration, swarm robotics, social dynamics, and animal behavior. A defining characteristic of these systems is the emergence of large...

📖 Read original article


34. Transforms for LLM Quantization: The Great Inversion and Format Co-Design ​

Author: Ehsan Jokar
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT

arXiv:2608.25188v1 Announce Type: new Abstract: Most competitive 4-bit LLM research pipelines now open the same way: apply a linear, function-preserving transform (rotation, scaling, permutation, non-orthogonal affine) so the outlier mass sits more favorably against the group scales, and only then r...

📖 Read original article


35. What Should a Large Language Model See? Physical Invariants as a Data Representation for PDE Discovery ​

Author: Fan Yang, Matt Thomson
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.soft

arXiv:2608.25189v1 Announce Type: new Abstract: Understanding how molecular interactions govern macroscopic behaviour is a central challenge in molecular sciences. However, conventional theory building cannot keep pace with the vast datasets modern experimentation routinely produces. Large language ...

📖 Read original article


36. Hyperbolic Latent Geometry for Tree-Structured Prototype Networks: A Local-vs-Global Trade-off ​

Author: Peter Flo, Luca Grossmann
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25199v1 Announce Type: new Abstract: We study a tree-structured regularizer over class-prototype layouts in a hierarchical-classification model and ask whether the choice of latent manifold for the prototypes (Euclidean R^d vs. the Poincare ball B^d_c) affects how well that regularizer ca...

📖 Read original article


37. Learning Mixtures of Plackett-Luce Models for Multi-Objective Alignment ​

Author: Dongyue Li, Ziniu Zhang, Lu Wang, Hongyang R. Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.25200v1 Announce Type: new Abstract: We consider the problem of learning a mixture of $k$ Plackett-Luce models given multi-way ranking responses from annotators that may represent heterogeneous underlying preferences. This problem has many applications in AI alignment and preference optim...

📖 Read original article


38. LibriBrain100: One Hundred Hours of Broad and Deep MEG Data for Neural Speech Decoding at Scale ​

Author: Francesco Mantegna, Dulhan Jayalath, Gereon Elvers, Tasha Kim, Benjamin Ballyk, Alex Fung, SungJun Cho, Teyun Kwon, Luisa Kurth, Miran "Ozdogan, Gilad Landau, Pratik Somaiya, Natalie Voets, Mark Woolrich, Oiwi Parker Jones
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.25204v1 Announce Type: new Abstract: We introduce LibriBrain100, a large-scale MEG dataset for speech decoding designed from the ground up for reproducible, standardised evaluation. LibriBrain100 more than doubles the size of the original LibriBrain release, resulting in over 100 hours of...

📖 Read original article


39. Representing MAX functions using two-hidden-layer ReLU networks ​

Author: Zhimao Wang, Amitabh Basu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2608.25221v1 Announce Type: new Abstract: We study exact representations of $\mathrm{MAX}_N(x)=\max{x_1,\ldots,x_N}$ using two-hidden-layer ReLU neural networks. This problem has been studied in recent years in an attempt to characterize the exact number of hidden layers required to represent ...

📖 Read original article


40. Trust the Mass: Forced Weights in KV-Cache Eviction ​

Author: Jack Shi, Jerry Gu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.25230v1 Announce Type: new Abstract: Every deployed sparse-attention or KV-cache-eviction rule keeps a subset of the keys, discards the rest, and renormalizes the attention weights over the kept set. Enumerating the exact best subset under that constraint on $168{,}192$ attention rows fro...

📖 Read original article


41. Output Dilution: Redundant but Fragile Representations in MoE Models ​

Author: Orion Reblitz-Richardson
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.25231v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models appear to encode moral content as robustly as dense models, yet prove far more fragile in their encoding. In OLMoE-1B-7B, linear probes recover moral valence from nearly every expert-layer combination, with mean peak-lay...

📖 Read original article


42. Long-Term Behavioral Evaluation for Trusted Collaborator Selection via Bidirectional Mamba ​

Author: Botao Zhu, Xianbin Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2608.25232v1 Announce Type: new Abstract: Effective selection of trustworthy collaborators is crucial to ensuring the successful completion of collaborative tasks, which requires accurate assessments of both long-term device behavior and short-term collaborative dynamics. Consistent device beh...

📖 Read original article


43. ShuttleArena: Interpretable Self-Play in Physics-Based Badminton ​

Author: Peize Ding
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25246v1 Announce Type: new Abstract: Badminton is a compact but challenging domain for game AI: a player must choose a physically feasible shuttle trajectory, anticipate the opponent's interception, and recover to a court position whose value depends on the opponent's next response. The c...

📖 Read original article


44. Neural-Bayesian Structure Learning for Discrete Choice Modeling ​

Author: Hyunsoo Yun, Eun Hak Lee, Jiaru Zhang, Ziran Wang, Eui-Jin Kim
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25258v1 Announce Type: new Abstract: Conventional discrete choice and machine learning models are estimated primarily from observational data and typically treat explanatory covariates as parallel inputs, providing no internal mechanism for determining how related attributes should adjust...

📖 Read original article


45. Mitigating LLM sycophancy with RL-based fine-tuning: Bayesian Truth Serum approach ​

Author: Serhii Mytsyk, Yiming Zhang, Vikram Krishnamurthy
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2608.25267v1 Announce Type: new Abstract: Large language models (LLMs) frequently exhibit \emph{sycophancy}: they adapt their answers to a user's stated beliefs or preferences instead of reporting what they hold to be true, which lowers factual accuracy and can amplify misinformation. This pap...

📖 Read original article


46. SHSP: Structure-Aware Hierarchical Solution Prediction for Mixed-Integer Linear Programming ​

Author: Zherong Zhang, Guanlin Li, Chengrui Gao, Haopu Shang, Ke Xue, Jixiang Lu, Weiyong Yang, Chao Qian
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25282v1 Announce Type: new Abstract: Mixed-Integer Linear Programming (MILP) is a fundamental optimization paradigm in combinatorial optimization and has been widely applied across real-world domains. Due to its NP-hard nature, obtaining optimal solutions for large-scale or highly constra...

📖 Read original article


47. InsightSR: Refining Symbolic Regression Search Spaces via Parallel Semantic and Structural LLM Guidance ​

Author: Yating Ling, Wenjing Cun, Zhitang Chen
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25291v1 Announce Type: new Abstract: Symbolic regression (SR) seeks to discover parsimonious mathematical laws from observational data, yet conventional approaches often struggle with the vast combinatorial search space of physically meaningful expressions. We present InsightSR, a framewo...

📖 Read original article


48. Prefix-Denoising Consistency: Test-Time Verification for Diffusion Language Models ​

Author: Yuki Ichihara, Naoto Iwase, Mohammad Atif Quamar, Junpei Komiyama
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25311v1 Announce Type: new Abstract: Diffusion Language Models (DLMs) have recently become increasingly competitive with autoregressive (AR) models, and even outperform them on certain tasks. Unlike AR models, DLMs produce output through iterative denoising without a left-to-right order. ...

📖 Read original article


49. Activation-Space Order-Swap Geometry: A Site-Asymmetry Audit ​

Author: Anqi Peter Li
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25315v1 Announce Type: new Abstract: Order-dependent activation statistics are often interpreted as evidence of interaction, but that interpretation can be confounded by where interventions enter the network. We introduce a no-fit site-asymmetry audit. For a twice-differentiable readout, ...

📖 Read original article


50. Two Dimensions Govern Agnostic Multiclass Transductive Learning ​

Author: Pahan Dewasurendra
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25326v1 Announce Type: new Abstract: In transductive classification, an adversary fixes a labeled population, one label is hidden uniformly, and the learner sees all remaining labels. For binary classes, agnostic transductive and PAC learning have the same minimax rate. Whether this exten...

📖 Read original article


51. Neither Precision Nor Architecture Alone: Controlled Tests of Failure Remedies for Physics-Informed Neural Networks ​

Author: Jinyuan Zhang, Peng He, He Hu, Yin Yuan, ShengShuo Jiao
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25327v1 Announce Type: new Abstract: Physics-Informed Neural Networks (PINNs) frequently fail on stiff or advection-dominated PDEs, and two recent accounts offer competing remedies: switching from FP32 to FP64 to repair an L-BFGS stopping artifact, or replacing the MLP with a state-space-...

📖 Read original article


52. Beyond Pairwise Feedback: Listwise Vision-Language Supervision for Preference-Based Reward Learning ​

Author: Srivalli Katkuri, Maxwell Kawada, Juan Wachs
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.25350v1 Announce Type: new Abstract: Vision-language models (VLMs) have emerged as a powerful source of supervision for reinforcement learning, enabling agents to leverage rich semantic knowledge during training. Inspired by the success of preference-based reward learning (PbRL) in reinfo...

📖 Read original article


53. Escaping Low-Dimensional Overlap: Multi-Task Model Merging via High-Dimensional Sparse Disentanglement ​

Author: Yihang Zhang, Shengke Sun, Junjie Wen, Feng Zeng
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.25354v1 Announce Type: new Abstract: Model merging provides an efficient way to construct multi-task generalist models without additional training, but its performance often degrades under severe task interference. Task interference in model merging primarily stems from \textit{superposit...

📖 Read original article


54. PaSta: Noisy Node Classification with Partial Label Learning ​

Author: Yujing Liu, Yixin Liu, Yu Zheng, Yue Tan, Alan Wee-Chung Liew, Shirui Pan
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25365v1 Announce Type: new Abstract: Noisy node classification problem is a fundamental yet challenging task for real-world graph-related web services, where node labels are often corrupted or unreliable due to weak supervision or automatic annotation. However, existing methods typically ...

📖 Read original article


55. Refusal geometry reflects refusal training: diverse refusal prefixes can raise stable rank and weaken refusal vector ablation attacks ​

Author: Andrey Labunets
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2608.25390v1 Announce Type: new Abstract: Refusal training protects AI models from jailbreaks by training models to decline unsafe queries, reducing the risk of misuse. Recent work finds that refusal behavior in aligned language models can be mediated by a single activation direction or a low-...

📖 Read original article


56. Joint Initialization of Flux Networks and Effective Multiplication Factor for Physics-Informed Neural Networks Solving Neutron Diffusion Problems ​

Author: Qin Hang, Yangdi Yi, Jiayi Li, Xu Wang, Heng Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25443v1 Announce Type: new Abstract: Efficient determination of the effective multiplication factor (keff) is an important computational task in reactor core neutronics analysis. Physics-informed neural networks (PINNs) incorporate neutron diffusion equations and boundary conditions into ...

📖 Read original article


57. Resolving Multi-Modal Regression by Difference-Quotient-Based Clustering:Fast Coarse Conditional-Label Assignment ​

Author: Huang Weiquan
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25467v1 Announce Type: new Abstract: Multimodal regression suffers from the mean-collapse pathology: under squared loss, an unconstrained regressor converges to the conditional mean, which for K > 1 lies away from all modes. We attribute this failure to pairwise contradictions--samples wi...

📖 Read original article


58. A Storage-Retrieval Gap in Parametric Knowledge Graph Memory ​

Author: Martino M. L. Pulici, Cuong Xuan Chu, Evgeny Kharlamov, Volker Tresp
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IR

arXiv:2608.25489v1 Announce Type: new Abstract: Graph retrieval-augmented generation places retrieved subgraphs into the model's context window at query time, paying a recurring token cost and exposing source data on every call. We study an alternative: compiling a knowledge graph offline into a ban...

📖 Read original article


59. FedQoS: Federated QoS-Risk Learning for Heterogeneous Indoor-Outdoor Access Selection ​

Author: Nguyen Van Thieu, Ti Ti Nguyen, Ons Aouedi, Zerihun Huruy, Vu Nguyen Ha, Symeon Chatzinotas
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25496v1 Announce Type: new Abstract: Reliable access selection in dynamic and heterogeneous indoor-outdoor environments is challenging because instantaneous radio measurements alone cannot capture future QoS degradation caused by mobility, blockage, traffic load, and resource competition....

📖 Read original article


60. Resilient Decentralized Wireless Federated Learning via Gradient Tracking with AdamW ​

Author: Nguyen Van Thieu, Ti Ti Nguyen, Ons Aouedi, Vu Nguyen Ha, Symeon Chatzinotas
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.NI

arXiv:2608.25535v1 Announce Type: new Abstract: Wireless Internet-of-Things (IoT) edge networks require decentralized learning (DecL) methods that can operate reliably under both heterogeneous local data and communication-constrained wireless links. However, existing decentralized optimization schem...

📖 Read original article


61. Reflection Steering: Disentangling Reflection from Reasoning in Activation Space for Token-Efficient Inference ​

Author: Jiarui Hu, Zhiyuan Wen, Xiaoyun Liu, Jiaxing Shen, Yu Yang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.25542v1 Announce Type: new Abstract: Large reasoning models often produce reasoning traces with verification, revision, and backtracking. When reflection merely re-checks established results, it wastes reasoning tokens and increases latency. Most existing reflection steering methods add a...

📖 Read original article


62. Interpreting Protein Language Model Embeddings via Orthogonal Projection for Protein Fitness Prediction ​

Author: Paulo Yanez Sarmiento, Pia Francesca Rissom, Manuel Pfeuffer, Marco Simnacher, Jordan F. Safer, Sumaiya Iqbal, Henrike O. Heyne, Nadja Klein, Bernhard Y. Renard
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM

arXiv:2608.25548v1 Announce Type: new Abstract: Recently, there has been a growing adoption of protein language models (PLMs) in biomedical science. Their embeddings provide a rich numerical representation of protein sequences which achieve state-of-the-art performance on several downstream tasks in...

📖 Read original article


63. Beyond Optimal Rates in Stochastic Optimization: Trajectory-Adaptive Stopping Rules ​

Author: Liviu Aolaritei, Lucas L'evy, Francis Bach, Michael I. Jordan
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, math.OC, math.ST, stat.ML, stat.TH

arXiv:2608.25551v1 Announce Type: new Abstract: Stochastic gradient descent (SGD) is typically analyzed at a deterministic horizon chosen before the algorithm is run, even though practical stopping decisions are made adaptively by inspecting the evolving trajectory. This mismatch creates a fundament...

📖 Read original article


64. Physics-Informed Foresight Pruning for Sparse PINN Solvers of Nonlinear PDEs ​

Author: Ahmad Ishaque Karimi, Uvini Balasuriya Mudiyanselage, Kookjin Lee
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25564v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) often rely on over-parameterized models to optimize coupled solution and differential-residual objectives, leaving unclear how much capacity is necessary and what pruning should preserve. We study foresight prun...

📖 Read original article


65. Beyond Scaling: Self-Evolving LLM Agents for Hardware Kernel Optimization via an Experience-Driven Workflow and Experience Graph Memory ​

Author: Siyuan Chen, Runlin Hou, Shenxiu Wu, Yansong Sun, Junming Cao, Yiyu Zhang, Shudi Shao, Junhao Qiu, Zhichao Lu, Qingfu Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.MA

arXiv:2608.25570v1 Announce Type: new Abstract: Hardware kernel optimization requires repeated compilation, correctness testing, profiling, and revision. LLM agents can automate parts of this process, and stronger foundation models, longer context windows, and longer execution horizons have improved...

📖 Read original article


66. Individual Fairness in Hierarchical Clustering ​

Author: Binita Maity, Shrutimoy Das
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25586v1 Announce Type: new Abstract: Hierarchical clustering produces ultrametric representations that impose strong global geometric constraints and may distort local similarities in ways that disproportionately affect individual data points. We study hierarchical clustering under an ind...

📖 Read original article


67. M-Fibration Theory with Applications to Neural Network Compression ​

Author: Paolo Boldi
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25598v1 Announce Type: new Abstract: The purpose of this paper is to provide a general, comprehensive, theoretical framework that allows one to deal with fibrations on graphs labelled on a commutative monoid. This is a genuine extension of the theory of graph fibrations (as introduced in ...

📖 Read original article


68. Frequency-aware forecasting for short-term typhoon gust prediction ​

Author: Xuefei Wang, Tingyi Liu, Heng Zhang, Shengjun Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25604v1 Announce Type: new Abstract: Accurate gust forecasting under typhoon conditions remains challenging due to the highly non-stationary and multi-scale characteristics of extreme wind fluctuations. Existing deep learning models often struggle to simultaneously capture long-term trend...

📖 Read original article


Author: Junzhao Zhang, Tao Zhang, Liren Yu, Feiyi Dong, Zhixuan Zhang, Dan Ou, Haihong Tang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2608.25635v1 Announce Type: new Abstract: Industrial e-commerce search systems ultimately aim to optimize the user-level long-term objective, such as n-day cumulative purchases or gross merchandise value (GMV) per user. However, such objectives are defined at the user level, whereas search ran...

📖 Read original article


70. A Token-Level Analysis of Sampled-Token Reverse-KL On-Policy Distillation ​

Author: Bing Shao, Jiazheng Zhang, Long Ma, Yujiong Shen, Senjie Jin, Xin Guo, Yuming Yang, Mingxu Chai, Zhiheng Xi, Tao Gui, Qi Zhang, Xuanjing Huang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.25643v1 Announce Type: new Abstract: On-policy distillation (OPD) supervises a student on its own trajectories with token-level signals from a frozen teacher, yet how a sampled loss allocates updates across tokens remains poorly understood. We analyze the gradient of the per-token K2 esti...

📖 Read original article


71. LDAC-Net: A Learnable Multi-Lag Differencing Attention-Convolution Network for Drift-Robust Recognition with Low-Cost MOX Gas Sensors ​

Author: Xin Zhang, Liangxiu Han, Yue Shi, Tam Sobeih
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25646v1 Announce Type: new Abstract: Portable electronic-nose systems based on low-cost metal-oxide (MOX) gas sensors offer a practical solution for gas and odour recognition, but their signals are affected by slow chemical transients, drifting sensor offsets, scale variation, and cross-c...

📖 Read original article


72. Adversarial Training of Linear Models under Stealthy Attacks ​

Author: Lovisa Eriksson, Dave Zachariah, Andr'e M. H. Teixeira
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.SY, eess.SY

arXiv:2608.25681v1 Announce Type: new Abstract: Predictive models are widely used in many fields, but are vulnerable to false data injection attacks. To address this, detection schemes and adversarial training have been proposed, but such approaches lack guarantees against stealthy attacks. We there...

📖 Read original article


73. Modeling spatio-temporal locality in multi-step forecasting of geo-referenced time series ​

Author: Annunziata D'Aversa, Gianvito Pio, Michelangelo Ceci
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25698v1 Announce Type: new Abstract: Forecasting future measurements from geographically distributed sensors is essential across many domains. However, the spatial distribution of these sensors raises multiple challenges, primarily due to spatial autocorrelation phenomena, that introduce ...

📖 Read original article


74. Tropospheric temperature and humidity profile retrieval from Meteosat Flexible Combined Imager based on deep learning ​

Author: Alejandro Salgueiro, Johannes Rausch, Julie Th'er`ese Villinger, Angela Meyer
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2608.25700v1 Announce Type: new Abstract: The Meteosat Third Generation (MTG) Flexible Combined Imager (FCI) offers new opportunities for tropospheric temperature and humidity profiling, at higher spatio-temporal resolutions and expanded spectral coverage relative to its predecessor. Verticall...

📖 Read original article


75. Fairness-Aware Test-Time Prompt Tuning ​

Author: Yoann Launay, Parameswaran Kamalaruban, Tom Kempton, Stuart Burrell, David Sutton
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25707v1 Announce Type: new Abstract: Vision-language models have displayed remarkable capabilities in multi-modal understanding and are increasingly used in critical applications where economic and practical deployment constraints prohibit re-training or fine-tuning. However, these models...

📖 Read original article


76. It's a matter of timescale: non-linear utility in successor features and multi-objective planning and learning ​

Author: Liam P. H. Mertens, Lucas N. Alegre, Florent Delgrange, Diederik M. Roijers, Ann Now'e, Peter Vamplew
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25723v1 Announce Type: new Abstract: Time is of the essence when dealing with multiple reward signals and non-linear utility. In this paper we argue that the current main approaches in multi-objectiveRL (SER and ESR), and successor features, are insufficient. While each approach deals wit...

📖 Read original article


77. Are LLM-Enhanced GNNs Privacy-Safe? ​

Author: Longzhu He, Zelang Wen, Chaozhuo Li, Sen Su
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.25727v1 Announce Type: new Abstract: Large language models (LLMs) have recently advanced graph neural networks (GNNs) by enriching node representations with semantic information, giving rise to LLM-enhanced GNNs that achieve substantial performance gains. However, their vulnerability to p...

📖 Read original article


78. Why Does Graph Learning Fail to Fully Benefit from a Text Teacher? ​

Author: Fumiaki Kimino (SOKENDAI), Ryoma Sato (SOKENDAI, National Institute of Informatics)
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.25741v1 Announce Type: new Abstract: Graph neural networks (GNNs) are widely used to represent complex interactions and relationships among entities. We investigate a multimodal model that combines two complementary ideas: a self-supervised method that enables a GNN encoder pretrained on ...

📖 Read original article


79. A Constitutive Markov Physics-Informed Neural Operator (MPNO) for Autoregressive Stability in Transient Dynamics ​

Author: Wenpu Du, Peng Zhou, Yunlong Xia, Sinuo Xin, Congcong Zhang, Boyang Zhang, Yi Zhang, Wenzheng Xu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25744v1 Announce Type: new Abstract: Neural operators applied to transient-dynamics PDEs with strong discontinuities exhibit autoregressive instability: in concrete-penetration stress-field prediction, the wavelet neural operator (WNO) diverges in autoregressive rollout, while MeshGraphNe...

📖 Read original article


80. Comparing Corrupted Constrained Learning Problems ​

Author: Laura Iacovissi, Rabanus Derr, Robert C. Williamson
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH

arXiv:2608.25745v1 Announce Type: new Abstract: A key result in statistics is the data processing inequality, originally proved by Blackwell (1951) and later refined by DeGroot (1962) in terms of statistical uncertainty. It states that the Bayes risk of a statistical experiment obtained by stochasti...

📖 Read original article


81. TailSFT: Filtered Fine-Tuning Improves Post-Training Performance ​

Author: Sadhika Malladi, Samy Jelassi, Dylan Foster, Jordan T. Ash, Akshay Krishnamurthy
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25756v1 Announce Type: new Abstract: Reinforcement learning post-training drives reasoning and agentic capabilities in modern AI systems, yet a growing body of work shows that it is most effective when used to fine-tune an already capable base model. We question whether existing pipelines...

📖 Read original article


82. Learning from waste: Machine Learning for health risk prediction and computer vision-based sorting in Ghana ​

Author: Hilda Adwubi Osei, Catherine Tenewaa Osei, Desdemona Yaa Asobayire
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.CY

arXiv:2608.25759v1 Announce Type: new Abstract: The inappropriate disposal of solid waste remains a significant public health and environmental concern worldwide, including in Ghana. Poor sanitation and improper waste management practices contribute to substantial economic costs and avoidable deaths...

📖 Read original article


83. Drift-Aware Multimodal User Representation Learning via Multi-Scale Temporal Modeling and Sparse Mixture-of-Experts ​

Author: Ziqing Qian, Haohang Chen, Shengqi Dang, Yuhan Xiong, Canyu Shen, Jiaying Lei, Nan Cao
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25773v1 Announce Type: new Abstract: Understanding user preferences from noisy and temporally evolving social media behaviors is fundamentally challenging due to interest drift, where user preferences shift across time and exhibit both multi-scale temporal patterns and diverse co-existing...

📖 Read original article


84. EXAONE Tabular 1.0 : Technical Report ​

Author: Moonjung Eo, Min-Kook Suh, Hye-Seung Cho, Jiwon Kim, Seoyoon Kim, Sangjun Nam, Soonyoung Lee
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25774v1 Announce Type: new Abstract: EXAONE Tabular is a compact tabular foundation model family for classification and regression via in-context learning, producing predictions without dataset-specific gradient updates. Pretrained exclusively on a synthetic structural-causal-model (SCM) ...

📖 Read original article


85. Cooperative Multi-Agent Reinforcement Learning for Adaptive Aggregation in Semi-Supervised Federated Learning with non-IID Data ​

Author: Rene Glitza, Luca Becker, Rainer Martin
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.SD, eess.AS, eess.SP

arXiv:2608.25794v1 Announce Type: new Abstract: Federated Learning (FL) enables distributed training of machine learning models while preserving data privacy. However, FL struggles with heterogeneous, non-IID client data distributions, resulting in sub-optimal and biased global models. In this paper...

📖 Read original article


86. Geometry-Constrained Kolmogorov-Arnold Networks: Learning Edge Geometry via Banach Duality ​

Author: K S Sesh Kumar
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.25807v1 Announce Type: new Abstract: Kolmogorov-Arnold Networks (KANs) replace fixed activations in deep architectures with learnable univariate edge functions, making the choice of edge parametrisation central. Existing variants rely on fixed bases such as splines, polynomials, or Fourie...

📖 Read original article


87. Canalization Before Generalization: Grokking as a Dynamical Probe ​

Author: Yiming Lin
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25813v1 Announce Type: new Abstract: For overparameterized neural networks, many solutions can fit the training data equally well while behaving very differently on unseen samples. Grokking separates training fit from visible generalization, providing a window for studying how this select...

📖 Read original article


88. Learning Continuous Regional Temperature Fields with Lead-Time and Resolution Queries ​

Author: Chunlei Shi, Jiong Wang, Yi-Lin Wei, Junming Hou, Jinjin Liu, Yecheng Zhang, Dan Niu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.MM

arXiv:2608.25823v1 Announce Type: new Abstract: Accurate regional near-surface temperature forecasting is fundamental to short-range weather services and downstream risk assessment. Existing deep learning-based regional forecasters commonly produce a fixed set of future frames on a prescribed grid, ...

📖 Read original article


89. VINCENT: Validated Interaction Network for Cross-drug Explanation of Therapeutics ​

Author: Fan-Sheng Chuang, Xuchen Li, Yujing Bian, Kaixiong Zhou
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25841v1 Announce Type: new Abstract: Drug synergy prediction estimates whether two drugs produce a stronger joint effect than expected from their individual activities. For drug combination discovery, a single synergy score is often not enough: researchers also need to know which molecula...

📖 Read original article


90. CEDAR: Controlled and Event-Driven Demand Forecasting via Residual Decomposition ​

Author: Junjie Meng, Ranxu Zhang, Zi-an Zhang, Shujun Liu, Xiaoning Qi, Xiaozhou Xu, Yanyong Zhang, Hui Xiong, Chao Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25871v1 Announce Type: new Abstract: Forecasting in large-scale e-commerce marketplaces is increasingly required to support planning: merchants need to evaluate sales outcomes under future action sequences such as budget schedules, rather than passively predicting what happens next. Howev...

📖 Read original article


91. How Edge of Stability Hinders SCAFFOLD in Federated Optimization ​

Author: Anant Khandelwal, Michael Crawshaw, Mingrui Liu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25873v1 Announce Type: new Abstract: In federated learning, it is well known that heterogeneous data can (in theory) slow down optimization, and much effort has been directed at designing optimization algorithms that are unaffected by data heterogeneity, such as the SCAFFOLD algorithm. Ye...

📖 Read original article


92. A General-Purpose Molecular Foundation Model Transfers Across Diverse Olfactory Tasks ​

Author: Yikun Han, Yi Wang, Neil Mankodi, Stephen Yang, Ambuj Tewari
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25893v1 Announce Type: new Abstract: Foundation models have transformed molecular property prediction, yet it remains unclear whether a molecular foundation model, fine-tuned on a single canonical olfactory prediction task, can learn representations that transfer across diverse machine ol...

📖 Read original article


93. Towards A Unified Information Bottleneck Framework for Time Series Explanations ​

Author: Xu Zheng, Zichuan Liu, Zhuomin Chen, Mayur Akewar, Janki Bhimani, Jason Liu, Mo Sha, Jingchao Ni, Wei Cheng, Dongsheng Luo
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25897v1 Announce Type: new Abstract: Explaining deep learning models operating on time series data is crucial in various applications that require transparent and interpretable insights into model behavior. {Existing explanation methods generally fall into two categories: attribution-base...

📖 Read original article


94. Forecasting Multiple Observables with SCROLL: Score-Trained Uncertainty for Stochastic Dynamics ​

Author: Pavel Prochazka
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25898v1 Announce Type: new Abstract: Forecasting a stochastic dynamical system rarely means a single number: one wants several observables---future state, threshold event, regime label---each with its own likelihood. Standard multi-task recipes balance per-task losses, tuned or learned. W...

📖 Read original article


95. Quantum-Inspired Modeling of Driving Behavior ​

Author: Mohammad Elayan, Omid Armantalab, Wissam Kontar
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY, stat.ML

arXiv:2608.25907v1 Announce Type: new Abstract: Driver behavior is heterogeneous, context-dependent, and changes over time, and these properties shape the traffic phenomena we observe. Most models, however, fix in advance which behavioral variables interact and how. Behavior outside that form is abs...

📖 Read original article


96. One Symptom, Three Levers: A Critical Review of On-Policy Self-Distillation ​

Author: Justin Robert, Raheel Qader
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.25936v1 Announce Type: new Abstract: On-policy distillation trains a language model on its own generations while a teacher scores them token by token. It combines the dense supervision of imitation learning with the on-policy sampling of reinforcement learning. But it requires a second, l...

📖 Read original article


97. When Pruning Meets Interpretability: Preserving Sparse Autoencoder Robustness in LLMs ​

Author: Suchit Gupte, Xueru Zhang, Mohammad Mahdi Khalili
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25941v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) are widely used to interpret the internal representations of large language models (LLMs), yet their reliability under post-hoc model compression remains poorly understood. We present a systematic study of how pruning affects...

📖 Read original article


98. Spectral Allocation: Why Muon Outperforms Adam, and How to Improve Muon ​

Author: Xiaodong Wu, Wenyi Yu, Chao Zhang, Philip Woodland
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25990v1 Announce Type: new Abstract: Orthogonal optimisers such as Muon can substantially accelerate large language model pretraining relative to Adam, yet the mechanism remains incompletely understood. We investigate this through an out-of-sample spectral probing analysis of Transformer ...

📖 Read original article


99. DualOPSD: Adaptive Privileged Teachers for On-Policy Self-Distillation ​

Author: Yutong Chen, Guangfu Guo, Zhichao Xu, Kunpeng Liu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.26019v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) uses a privileged copy of the student model to provide dense supervision without an external teacher. OPSD keeps this privileged teacher fixed, even though the student distribution and output style change during train...

📖 Read original article


100. Robust CurveMoE: Multi-Norm Adversarial Defense for Mixture-of-Experts Models via Mode Connectivity ​

Author: Xu Zhang, Ren Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.26043v1 Announce Type: new Abstract: Multi-norm adversarial defense aims to protect neural networks against perturbations defined by different norm constraints, but existing methods typically optimize competing robustness objectives within a single parameter configuration, leading to subs...

📖 Read original article


101. How Much Rank Does LoRA Need? Rank-Error Bounds for Transformer Attention ​

Author: Gerard Conangla Planes
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.26052v1 Announce Type: new Abstract: Choosing the rank of a low-rank adaptation (LoRA) update is usually an empirical task. In this paper, we provide a task-dependent theory of the approximation error achievable at each LoRA rank for Transformer attention. We fix a pretrained attention he...

📖 Read original article


102. Group-Shared Low-Rank Approximation for Mobile-Efficient Pointwise Convolutions in Large-Kernel CNNs ​

Author: Hao Luo, Yiting Yang, Wenyi Zhao, Man Jiang, Zhijun Lin, Ghulam Mohiuddin, Ting Jiang, Kunming Luo, Zihao Zhang, Qingsen Yan, Guoqing Wang, Wei Dong, Peng Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.26069v1 Announce Type: new Abstract: Large-kernel Convolutional Neural Networks (CNNs) deliver remarkable performance in vision tasks by significantly expanding receptive fields, yet their quadratic parameter growth critically impedes storage-efficient edge deployment. While existing effi...

📖 Read original article


103. ICON Decomposition: Multivariate Concept-Level Explanations of Deep Representations for Model Auditing ​

Author: Roshan Prakash Rane, Marco Simnacher, Manuel Pfeuffer, Marc-Andre Schulz, Nys Tjade Siegel, Maximilian Dreyer, Frederik Pahde, Wojciech Samek, Sonja Greven, Kerstin Ritter
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, stat.ML

arXiv:2608.26083v1 Announce Type: new Abstract: Deep neural networks often exploit spurious associations in their training data, a failure known as shortcut learning. Concept-based explainability methods screen for shortcuts by testing whether concepts such as a patient's sex or scanner settings can...

📖 Read original article


104. TraceML: An Empirical Analysis of Human-Agent Planning in Machine Learning Development ​

Author: Jiarui Yan, Weiwei Sun, Sijie Li, Wenhan Li, Yiming Yang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.26086v1 Announce Type: new Abstract: Large language models write correct code for isolated problems but remain far weaker at autonomous machine-learning development, where an agent must revise data pipelines, models, and validation over hours of feedback, and on most competitions still fi...

📖 Read original article


105. Agentic Autoresearch for Cell-Edge Power Control: Radically Redefining the Researcher's Role ​

Author: Ahmad Khan, Akram Bin Sediq, Sara Azadegi Naeini, Raviraj S. Adve
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, cs.SY, eess.SY, math.IT

arXiv:2608.26093v1 Announce Type: new Abstract: Designing machine learning algorithms for wireless resource management is labour-intensive: the architecture, the loss function and the training recipe are all specified by hand. We demonstrate that this design layer can be surrendered to an autonomous...

📖 Read original article


106. Same-Player Verification for Account Consistency in Counter-Strike 2 ​

Author: Xuchen Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.24893v1 Announce Type: cross Abstract: In competitive first-person shooter (FPS) games such as Counter-Strike 2 (CS2), account-integrity review often asks whether an account's recent behavior remains consistent with its historical operator. This consistency question arises in cases such a...

📖 Read original article


107. Forecasting Weather-Driven Price Dynamics Across Sri Lankan Tea Market Catalogues ​

Author: Hesandi Mallawarachchi, Senilka Madurapperumage, Nadil Kulathunge, Thilokya Angeesa, Nethsith Gunaweera, Sandeepa Weerasekara, Patalee Narasinghe, Nisansa de Silva, Sandareka Wickramanayake
Published: 8/27/2026, 4:00:00 AM
Categories: econ.GN, cs.LG, q-fin.EC

arXiv:2608.24894v1 Announce Type: cross Abstract: The Colombo Tea Auction (CTA) plays a vital role in determining global tea prices, yet the relationship between local weather conditions and price behavior across different tea catalogues has not been thoroughly explored. In this study, we develop a ...

📖 Read original article


108. Detection != Reliable Control: Decodable Empathy Directions Yield at Most Partial Shifts in Automated Empathy Scores ​

Author: Haoran Jisun
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.HC, cs.LG

arXiv:2608.24901v1 Announce Type: cross Abstract: A decodable "empathy" direction is routinely read as a causal lever, conflating decodability, automated-metric control, and human-perceived change. We test this for two EPITOME-derived facets -- Recognition (cognitive) and Resonance (affective) -- in...

📖 Read original article


109. Evidence-Grounded Mapping of Multimodal Human Sensing Psychological Transdiagnostic Dimensions ​

Author: Xiyun Hu, Xiangyuan Xue, Yuting Lyu, Hanya Shao, Jingping Nie
Published: 8/27/2026, 4:00:00 AM
Categories: cs.HC, cs.LG

arXiv:2608.24903v1 Announce Type: cross Abstract: Mobile and wearable sensing enables longitudinal observation of behavior, yet translating these signals into meaningful mental health constructs remains difficult. We introduce a clinician-in-the-loop benchmark for evaluating whether large language m...

📖 Read original article


110. PA-CoT: Profile-Adaptive Chain-of-Thought for Personalized Nutritional Consulting ​

Author: Evgenii Garmashov, Nikita Kulin, Artur Khairullin, Viktor Zhuravlev, Daniil Sukhorukov, Mikhail Mozikov, Ilya Makarov, Sergey Muravyov
Published: 8/27/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.CL, cs.LG

arXiv:2608.24907v1 Announce Type: cross Abstract: In health and nutrition consulting, widely used prompting methods pass the user profile as an unstructured block without a dedicated analysis step, leaving personalization as a critical structural gap. We introduce PA-CoT (Profile-Adaptive Chain-of-T...

📖 Read original article


111. Modality Contribution Score - A Per-Patient Framework for Quantifying the Relative Diagnostic Contribution of Structural MRI and Amyloid PET in Alzheimer's Disease ​

Author: Dawa Chyophel Lepcha, Aaliya Ali, Sophie A. Martin, Deepika Koundal, Pierrick Coupe, Shabbir Syed-Abdul
Published: 8/27/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG

arXiv:2608.24931v1 Announce Type: cross Abstract: Multimodal neuroimaging combining structural MRI and positron emission tomography (PET) captures complementary structure-function relationships across the Alzheimer's disease (AD) continuum, yet existing artificial intelligence systems produce a sing...

📖 Read original article


112. The Dialect Tax: Dialectal Biases Persist throughout the Language Modeling Pipeline ​

Author: Elle
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.24952v1 Announce Type: cross Abstract: Systematic dialectal performance gaps in language models (LMs) are well documented, but the source of these disparities within the modern language modeling pipeline remains unclear. Our study traces this "dialect tax" across the natural language proc...

📖 Read original article


113. Beyond Tokens: Probing Higher-Order Epistasis in Learned Protein Representations ​

Author: Maryam Rahimimovassagh, Ivan Garibay, Niloofar Yousefi
Published: 8/27/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG

arXiv:2608.24953v1 Announce Type: cross Abstract: Protein fitness landscapes contain nonlinear interactions in which mutation effects depend on other residues. We introduce ORBIT, an Order-Resolved Benchmarking of Interaction Transformations framework that separates interaction presence, representat...

📖 Read original article


114. Common-Center Geometry and Certified Radial Reconstruction for Energy-Form Full Conformal Regions ​

Author: Yiheng Feng
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.24964v1 Announce Type: cross Abstract: This note studies the geometry of full conformal prediction (FullCP) regions generated by an empirical energy-form pairwise score. Candidate-score convexity alone does not guarantee connected FullCP regions, even when the candidate score is an empiri...

📖 Read original article


115. HCC+: Hyperbolic Guarding for Certified Attention Retrieval ​

Author: Liangchen Ge
Published: 8/27/2026, 4:00:00 AM
Categories: cs.DS, cs.LG

arXiv:2608.24971v1 Announce Type: cross Abstract: We study the Lipschitz stability of attention retrieval in hyperbolic spaces. Existing methods lack deterministic guarantees on attention-weight preservation under finite-precision representations. We introduce HCC+, a theoretical framework exploitin...

📖 Read original article


116. Retrieved But Not Reliable: A Survey on Attacks, and Defenses in Retrieval-Augmented Generation ​

Author: Minh Tran, Cuong Dang, Tuc Nguyen, Khanh-Tung Tran, Minh Huynh Nguyen, Trinh Chau, Kien Le, Do Xuan Long, Jiahao Zhang, Hoang D. Nguyen, Thanh Le, Suhang Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG

arXiv:2608.24977v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances large language models by grounding outputs in external knowledge, improving factuality and reducing hallucinations. At the same time, the retrieval-augmented pipeline introduces new robustness and securit...

📖 Read original article


117. Unsupervised Post-Training of Foundation Models: A Survey ​

Author: Yijie Xu, Qianyi Cai, Huizai Yao, Yili Wang, Tianfu Wang, Cehao Yang, Xingbo Yao, Zhiyu Guo, Aiwei Liu, Xuming Hu, Weiyu Guo, Hui Xiong
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CV, cs.LG, cs.MM

arXiv:2608.24982v1 Announce Type: cross Abstract: Foundation-model post-training usually relies on human labels, preference data, stronger teachers, or executable verifiers. We study Unsupervised Post-Training (UPT): update-bearing adaptation on unlabeled inputs whose learning signal is derived from...

📖 Read original article


118. Does Fine-Tuning Undo Activation Steering? Behavioural Recovery Without Weight-Edit Reversal ​

Author: Philipp E. Glass, Allan Tucker, Yongmin Li, Alina Miron
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.24988v1 Announce Type: cross Abstract: Activation steering can be embedded directly into a language model's weights, shaping behaviour without inference-time intervention and offering a way to encode alignment prior to release. However, models are routinely fine-tuned after deployment, an...

📖 Read original article


119. The Imperfective Paradox Is Not Necessarily in Large Language Models: A Benchmark Failure Before a Model Failure ​

Author: Kaiqiao Han, Yizhou Sun
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.25005v1 Announce Type: cross Abstract: The imperfective paradox provides a useful test of compositional semantic analysis. Recent work constructs an NLI benchmark and reports that models frequently infer completed telic events from progressive descriptions, attributing this behavior to a ...

📖 Read original article


120. EncoTESS: Age-Sensitive Encodings from Raw TESS Light Curves ​

Author: Phil R. Van-Lane (David A. Dunlap Department of Astronomy and Astrophysics, University of Toronto, Dunlap Institute for Astronomy and Astrophysics, University of Toronto, Department of Astronomy and Astrophysics, University of California San Diego), Joshua S. Speagle (David A. Dunlap Department of Astronomy and Astrophysics, University of Toronto, Department of Statistical Sciences, University of Toronto, Dunlap Institute for Astronomy and Astrophysics, University of Toronto, Data Sciences Institute, University of Toronto), Ryan Cloutier (Department of Physics and Astronomy, McMaster University), Christopher A. Theissen (Department of Astronomy and Astrophysics, University of California San Diego), Gwendolyn M. Eadie (David A. Dunlap Department of Astronomy and Astrophysics, University of Toronto, Department of Statistical Sciences, University of Toronto, Data Sciences Institute, University of Toronto), Ilay Kamai (Physics Department, Technion Israel Institute of Technology)
Published: 8/27/2026, 4:00:00 AM
Categories: astro-ph.SR, astro-ph.GA, astro-ph.IM, cs.LG

arXiv:2608.25019v1 Announce Type: cross Abstract: Main sequence stars of spectral types late F through M exhibit systematic variability in photometric light curves, particularly when they are young. Rotational modulation of starspots manifests as quasi-sinusoidal variability, which enables the measu...

📖 Read original article


121. Behind the [MASK]: Disentangling Representation and Faithfulness in DAPF-Based Dementia Detection ​

Author: Pardis Ranjbar-Noiey, Natalie Parde
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.25028v1 Announce Type: cross Abstract: Spoken-language analysis via prompt-based domain-adaptive models is a promising direction for low-resource, non-invasive dementia screening, but such models remain internally opaque. We study the interpretability of the Domain-Adapted models via Prom...

📖 Read original article


122. Improved Analysis for Hessian-free High-resolution Monte Carlo Sampling ​

Author: Wujun Lv, Xiaoyu Wang, Yingli Wang, Lingjiong Zhu
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.AP, math.PR

arXiv:2608.25052v1 Announce Type: cross Abstract: Hessian-free high-resolution (HFHR) dynamics augments underdamped Langevin dynamics (ULD) with reversible position diffusion for sampling problems that arise in machine learning. We establish an explicit quantitative contraction rate for HFHR dynamic...

📖 Read original article


123. DataKernelBench: Can LLMs Optimize Database Queries on GPUs? ​

Author: Gokul Karthik Kumar, Yotam Perlitz, Corey Lammie, Andrea Giovannini, Katja Hose
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.DB, cs.LG, cs.PL

arXiv:2608.25061v1 Announce Type: cross Abstract: GPUs increasingly accelerate database systems, but query-specific peak performance still often relies on hand-written kernels. Existing LLM kernel benchmarks focus on machine learning operators, leaving irregular, heterogeneous, data-movement-heavy d...

📖 Read original article


124. Scalable Self-Supervised Learning for Multiphase AC-OPF in Distribution Systems with Topology Reconfiguration ​

Author: Hoang T. Nguyen, Shaohui Liu, Reetam Sen Biswas, Varsha Pendyala, Nurali Virani, Deepjyoti Deka, Priya L. Donti
Published: 8/27/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC

arXiv:2608.25095v1 Announce Type: cross Abstract: The proliferation of distributed energy resources (DERs) in distribution grids enables the active coordination of these assets to reduce costs and enable cleaner operations. Realizing this potential requires solving multiphase AC optimal power flow (...

📖 Read original article


125. Towards Reliable, Generalizable, and Specific In-Context Knowledge Editing via Multi-Objective Reinforcement Learning ​

Author: Xuzhong Wang, Maiqi Jiang, Tejal Nair, Girija Bhusal, Yanfu Zhang, Haipeng Chen
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.25100v1 Announce Type: cross Abstract: Large Language Models (LLMs) are powerful but limited by static parametric knowledge that becomes outdated once pretraining ends. Knowledge editing addresses this problem by updating model behavior on target facts without full retraining. In particul...

📖 Read original article


126. FuzzingBrain-Bench V1: Evaluating Open-Ended Bug Discovery by LLMs ​

Author: Ze Sheng, Aleksandar Kezic, Zhicheng Chen, Jeff Huang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG, cs.SE

arXiv:2608.25158v1 Announce Type: cross Abstract: Evaluating the ability of large language models (LLMs) to discover software bugs is increasingly important. Existing benchmarks typically evaluate this capability by asking the model to generate a proof-of-concept input that triggers a predefined tar...

📖 Read original article


127. ROMNet: a hybrid reduced order modeling and machine learning approach to waveform inversion ​

Author: Liliana Borcea, Alexander Mamonov, Kui Ren, Haizhao Yang, Chugang Yi
Published: 8/27/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, physics.geo-ph

arXiv:2608.25160v1 Announce Type: cross Abstract: Waveform inversion seeks to estimate the wave speed of a heterogeneous, inaccessible medium, from time-resolved measurements of the waves at user controlled sensors. We consider this inverse problem for acoustic waves and an active array of source/re...

📖 Read original article


128. Lowering the Barrier to AI-Driven Inspection: A No-Code Workflow for Automated Structural Defect Detection ​

Author: Michael Holm, Tanner McElroy, Xinghang Zhang, Guang Lin
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV

arXiv:2608.25176v1 Announce Type: cross Abstract: Structural health monitoring (SHM) is essential in modern engineering, providing data for condition-based maintenance, lifecycle assessment, and predictive decision-making. Traditionally, SHM relied on visual inspection to detect defects such as crac...

📖 Read original article


129. Minimax Alternating Regret for the Experts Problem and Online Convex Optimization ​

Author: Mengxiao Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.25182v1 Announce Type: cross Abstract: In this paper, we study alternating regret in online convex optimization (OCO), motivated by the success of alternating learning dynamics in two-player games. Although previous works have shown that $o(\sqrt{T})$ alternating regret is achievable unde...

📖 Read original article


130. BanglaMamba: Exploring State Space Models for Bangla Fake News Detection ​

Author: M. K. Khalidi Siam
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.25190v1 Announce Type: cross Abstract: Fake news detection has become an important Natural Language Processing (NLP) task due to the rapid spread of misinformation through online news platforms and social media. While transformer-based models such as BanglaBERT achieve strong performance ...

📖 Read original article


131. Simulating Cognitive Smart Freight Corridors with Agent-Based Models and Reinforcement Learning ​

Author: Madelaine Martinez-Ferguson, Chun Wang, Mustafa Can Camur, Xueping Li
Published: 8/27/2026, 4:00:00 AM
Categories: cs.ET, cs.LG

arXiv:2608.25193v1 Announce Type: cross Abstract: Smart freight corridors offer a practical pathway for connected and automated vehicle (CAV) deployment in freight transportation, but physical experimentation is expensive and existing approaches rely on predefined control policies that cannot captur...

📖 Read original article


132. Multi-View Trust Evaluation for Collaborator Selection via Evidential Deep Learning ​

Author: Botao Zhu, Xianbin Wang
Published: 8/27/2026, 4:00:00 AM
Categories: eess.SY, cs.CR, cs.LG, cs.SY

arXiv:2608.25235v1 Announce Type: cross Abstract: Selection of trustworthy collaborators in distributed systems is critical for efficient task completion, necessitating the inference of trustworthiness from their past collaboration experience. However, as a collaborator serves distinct devices acros...

📖 Read original article


133. TrustFormer: Cross-Temporal and Cross- Dimensional Transformer for Task-Specific Multi-Dimensional Trust Evaluation ​

Author: Botao Zhu, Xianbin Wang
Published: 8/27/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2608.25238v1 Announce Type: cross Abstract: In dynamic collaborative systems, the selection of reliable collaborators is critical to ensuring effective task execution. Existing trust evaluation methods often rely on unidimensional or scalar representations, which fail to faithfully capture a c...

📖 Read original article


134. From Memorization to Absorption: Mixed-Policy RL for Continual Knowledge Injection ​

Author: Zhibo Hou, Fan Zhao, Zhiyu An, Wan Du
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.25243v1 Announce Type: cross Abstract: Continual knowledge injection is essential for keeping large language models up-to-date in a fast-evolving world. Existing methods rely on supervised fine-tuning (SFT), which memorizes injected facts in their training format but fails to generalize a...

📖 Read original article


135. Generative Action-Chunk Sampling for Adaptive Stiffness Control in Physical Human-Robot Collaboration ​

Author: Aoi Otake, Ferdinand Hartmann, Ko Igari, Shingo Murata
Published: 8/27/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2608.25284v1 Announce Type: cross Abstract: Physical human-robot collaboration requires a robot to provide assistance when human intention is clear while remaining compliant when several future motions are plausible. We present an adaptive stiffness framework based on generative action-chunk s...

📖 Read original article


136. WAVE: Reversing the Guidance Hierarchy for Coarse-to-Fine Guided Depth Super-Resolution ​

Author: Tayyab Nasir, Daochang Liu, Ajmal Mian
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.25302v1 Announce Type: cross Abstract: Guided depth super-resolution (GDSR) typically extracts RGB guidance features through convolutional hierarchies, inheriting their fine-to-coarse bias. Thus, low-level spatial cues surface in early layers, leaving the deeper layers to suppress those t...

📖 Read original article


137. SAUSS: Stochastic Approximation with Unbiased Simulated Scores for Limited Dependent Variable Models ​

Author: Sokbae Lee, Yuan Liao, Myung Hwan Seo, Youngki Shin
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, econ.EM

arXiv:2608.25304v1 Announce Type: cross Abstract: Multinomial choice models allow flexible substitution patterns but become computationally demanding with many alternatives or observations. With a fixed per-observation simulation budget, simulated maximum likelihood introduces simulation bias, while...

📖 Read original article


138. CRAMER: Control via Request-Aware Masking for Editing Recommenders ​

Author: Zhiyuan Julian Su, Naihe Feng, Zhen Luther Qin, Ga Wu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2608.25370v1 Announce Type: cross Abstract: Sequential recommendation models, while powerful, have limited flexibility in responding to immediate user requests, making it difficult to adapt their recommendations to the user's timely interests. Unfortunately, existing user request adaptation me...

📖 Read original article


139. A meta-algorithm for ab initio reconstruction of complex mixtures in cryo-EM ​

Author: Alkin Kaz, Arda Kaz, Ellen D. Zhong
Published: 8/27/2026, 4:00:00 AM
Categories: q-bio.BM, cs.LG

arXiv:2608.25388v1 Announce Type: cross Abstract: We describe a systematic approach for spawning and aggregating multi-class cryo-EM reconstruction jobs. This approach formalizes standard ad hoc strategies of iterative classification and filtering typically used by practitioners to sort impure, hete...

📖 Read original article


140. Token-Oriented Semantic Communication with Pretrained Vision Transformers ​

Author: Jiwoong Im, Minwoo Kim, Jaeho Lee, Yo-Seb Jeon, Yongjune Kim
Published: 8/27/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.CV, cs.LG

arXiv:2608.25410v1 Announce Type: cross Abstract: Token communications realize the semantic communication principle at the granularity of transformer tokens, providing a promising direction for client--server collaborative inference in resource-constrained edge systems. However, directly transmittin...

📖 Read original article


141. BVR Sim: An Open and High-Throughput Environment for Heterogeneous Air-Combat Reinforcement Learning ​

Author: Haocheng Sun (Beijing University of Posts,Telecommunications), Mulai Tan (Air Force Engineering University)
Published: 8/27/2026, 4:00:00 AM
Categories: cs.MA, cs.LG

arXiv:2608.25419v1 Announce Type: cross Abstract: Beyond-visual-range (BVR) air combat is a challenging reinforcement-learning domain characterized by partial observability, long-horizon decision making, energy management, and limited weapons. We present BVR Sim, an open-source Gymnasium-style envir...

📖 Read original article


142. Data-driven Effective Modeling of Stochastic Chemical Reaction Networks ​

Author: Yuan Chen, Weize Mao, Dongbin Xiu
Published: 8/27/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.DS

arXiv:2608.25421v1 Announce Type: cross Abstract: The Stochastic Simulation Algorithm (SSA), widely considered an exact algorithm for stochastic chemical reaction networks, suffers from high computational cost. In this work, we propose a data-driven effective model that operates on a user-defined co...

📖 Read original article


143. Distance Is Not Enough: Forget-Retain Alignment Gap Predicts LLM Relearning Robustness ​

Author: Yi Chen, Hanna Hsieh, Shuhong Liu, Chuanbo Hua, Zihan Ma, Kun Wang, Joo-Young Kim
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.25429v1 Announce Type: cross Abstract: Machine unlearning aims to make a model forget specific data, yet unlearned LLMs often fail to stay unlearned: brief fine-tuning can revive removed knowledge. Existing robustness predictors rely on global weight-space displacement, but distance alone...

📖 Read original article


144. Energy Yield and Lifetime Climate Classification via Machine Learning for Optimizing Photovoltaic Module Design and Materials ​

Author: Youri Blom, Sofia Dutto, Alexandru Costache, Rowan Richie, Ruben Pelsser, Wesley Berger, Jing Sun, Rudi Santbergen, Olindo Isabella, Malte Ruben Vogt
Published: 8/27/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG

arXiv:2608.25448v1 Announce Type: cross Abstract: To resiliently and sustainably meet our future energy demand, photovoltaic (PV) modules must be deployed across a broad and diverse range of geographical regions with varying operating conditions. As these conditions strongly affect both performance ...

📖 Read original article


145. Training Alignment Auditors via Reinforcement Learning ​

Author: Paul Rosu, Rowan Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.25460v1 Announce Type: cross Abstract: Alignment auditing of frontier models increasingly relies on LLM auditors to surface undesirable behaviors at scale, but current automated auditors can struggle with coherent investigation and audit realism. In this work, we improve LLM auditors with...

📖 Read original article


146. Functional linear regression from sparse to dense designs: a pooling-ridge method and minimax optimality ​

Author: Shunxing Yan, Fang Yao
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, math.ST, stat.ML, stat.TH

arXiv:2608.25468v1 Announce Type: cross Abstract: Functional data analysis is an important statistical field that treats data as random functions. In practice, the random functions are often not fully observed but instead measured at discrete times. While simpler problems, such as mean and covarianc...

📖 Read original article


147. AERIS: Offline Policy Improvement for Multi-UAV Integrated Sensing and Communication ​

Author: Ziyuan Wang (Steven), Yifan Sui (Steven), Wei Wei (Steven), Wenjie Xin (Steven), Zekai Zhang (Steven), Xiangwang Hou (Steven), Xiao-Ping (Steven), Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2608.25477v1 Announce Type: cross Abstract: Unmanned aerial vehicle (UAV)-enabled integrated sensing and communication (ISAC) is a promising 6G paradigm, but dynamic multi-UAV ISAC control must jointly balance communication quality, sensing reliability, and flight safety under stochastic mobil...

📖 Read original article


148. A Multi-View Coupled Tensor Decomposition for Lightweight Online Adaptive Traffic Prediction ​

Author: Quan Yu, Jie Ni, Yu-Hong Dai, Xiongjun Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2608.25498v1 Announce Type: cross Abstract: Accurate online traffic prediction is essential for intelligent transportation systems, where forecasting must be performed continuously under imperfect sensing conditions. Missing observations and anomalous disturbances make this task challenging, p...

📖 Read original article


149. Adaptive Regularization for Random Features: A Neighboring Early-Stopping Rule with Oracle-Rate Guarantees ​

Author: Caixing Wang, Zhibo Chen, Yue Wang
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.25513v1 Announce Type: cross Abstract: Random feature methods provide a scalable approximation to kernel ridge regression (KRR), but the regularization parameter that yields the oracle learning rate depends on unknown smoothness and capacity parameters. In this work, we propose a neighbor...

📖 Read original article


150. Adaptive Hybrid Subspace Levenberg Marquardt Algorithm with Adequacy Monitor for Large Scale Least Squares Problems ​

Author: M. Duc Hoang, Timothy J. Lewis
Published: 8/27/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.OC

arXiv:2608.25524v1 Announce Type: cross Abstract: The Levenberg-Marquardt (LM) algorithm is the most widely used method for solving nonlinear least-squares problems, as it combines the robustness of steepest descent with the fast local convergence of the Gauss-Newton method. However, its computation...

📖 Read original article


151. CropCop: An Auditable 120-Class Plant-Health Model from Benchmark Reconstruction to a Quantised Runtime Artifact ​

Author: Rana Muhammad Ahmed, Sabahat Abbas
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.25539v1 Announce Type: cross Abstract: A plant-health score can appear precise while resting on duplicated image families, a long-tailed label space, or a runtime file that was never evaluated. We present CropCop, a closed-set recognition system spanning 120 operational plant-health class...

📖 Read original article


152. Virgil: Navigating Explainability for Transformer-based Language Models ​

Author: Martino Ciaperoni, Sezer Kutluk, Benedetta Muscato, Marta Marchiori Manerba, Fosca Giannotti
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.25555v1 Announce Type: cross Abstract: Explainability for transformer-based language models is becoming crucial as these systems are deployed in high-stakes applications. As a result, the ecosystem of explainability tools is rapidly evolving, becoming richer, but also more fragmented and ...

📖 Read original article


153. A Hierarchical Synergistic Deep Learning Framework Integrating Composition, Structure, and Ionic Transport for Solid-State Electrolyte Discovery ​

Author: Hongwei Du, Dingyang Lv, Baole Wei, Yongheng Li, Feng Yu, Ziheng Lu, Siqi Shi, Hong Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.AI, cs.LG

arXiv:2608.25592v1 Announce Type: cross Abstract: Inorganic solid-state electrolytes must combine high room-temperature ionic conductivity, a wide electrochemical window, excellent electronic insulation, and favorable mechanical compliance. Single models struggle to support reliable multi-objective ...

📖 Read original article


154. JIT-Agent: Scaling Harness Intelligence via Just-in-Time Harness Evolution ​

Author: Guibin Zhang, Leo Lu, Fangzhou Xie, Kang Zhu, Junhao Wang, Zhifei Xie, Zhaochen Yu, Zihang Liu, Zhongxiang Sun, Qiankun Li, Yue Liao, Heng Chang, Xiaobin Hu, Qibing Ren, Wangchunshu Zhou, Shuicheng Yan
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.25593v1 Announce Type: cross Abstract: Agent capability is not determined by the model alone. The agent harness, encompassing memory management, planning strategy, action protocol, and tool/skill orchestration, can dominate the contribution of the underlying foundation model. Yet harness ...

📖 Read original article


155. Narcissus: Program Synthesis Using Context-Aware LLM Approximations ​

Author: Tilman Hinnerichs, Sebastijan Dumancic, Neil Yorke-Smith
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.PL, cs.SE

arXiv:2608.25657v1 Announce Type: cross Abstract: Large language models (LLMs) excel at programming, but not when the task fixes the target language: prompted with a grammar rare in their training data, their programs usually break the grammar or fail the given specification. Enumerative synthesizer...

📖 Read original article


156. Learning New Facts with QLoRA: An Acquisition-Retention Frontier ​

Author: Estelle Zheng, S'ebastien Warichet, Emmanuel Helbert, Christophe Cerisara
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.25677v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning is often assumed to preserve pretrained capabilities because it updates only a small number of parameters. We show that this assumption depends strongly on adapter capacity. We study factual acquisition in a controlled...

📖 Read original article


157. Unsupervised Anatomical Feature Learning via Diffusion Models: Enhanced Medical Image Segmentation with Denoising Diffusion Probabilistic Models ​

Author: Akshat G, Divyansh Gupta, Shaleen Bhatnagar, Shilpa Ankalaki, Tusar Kanti Mishra
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.25693v1 Announce Type: cross Abstract: Acquiring pixel-level annotations for medical image segmentation is a severe bottleneck. Traditional U-Net architectures, while effective, learn local texture patterns and lack awareness of global anatomical structures, leading to boundary delineatio...

📖 Read original article


158. Fast rates in Bayesian online learning with approximate posteriors ​

Author: Ilsang Ohn
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.25706v1 Announce Type: cross Abstract: Exact Bayes prediction enjoys fast predictive regret guarantees, but exact posterior updating or representation may be too costly for online use. We study when these statistical guarantees are preserved by computational approximations. We show that t...

📖 Read original article


159. Multi-output Gaussian process prediction of physical fields under linear equality constraints ​

Author: Mahamat Hamdan Nassouradine, Cl'ement Gauchy, Pierre-Emmanuel Angeli, S'ebastien da Veiga
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.25709v1 Announce Type: cross Abstract: We address the simultaneous prediction of multiple high-dimensional physical fields governed by linear equality constraints, a setting that arises in many real-world applications in physics machine learning. Gaussian process (GP) regression is a wide...

📖 Read original article


160. Pointing the Way, Hiding the Destination: Practical Private Dense Retrieval at Scale ​

Author: Peichun Hua, Danyang Chen, Junan Zhang, Haifeng Sun, Jingyu Wang, Diwen Xue, Mingyu Li, Yunming Xiao
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.IR, cs.LG

arXiv:2608.25735v1 Announce Type: cross Abstract: Hosted retrieval-augmented generation (RAG) and semantic search allow users to query valuable provider-held corpora, raising two competing demands: to hide each query and chosen result, yet reveal only the documents that the user is authorized to rec...

📖 Read original article


161. MeMark: Membrane-Space Watermarking for Spiking Neural Networks ​

Author: Roberto Ria~no, Gorka Abad, Stjepan Picek, Aitor Urbieta
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.25738v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) are increasingly distributed as pretrained checkpoints and reused as backbones for new tasks. However, current SNN watermarks are mainly verified against the model output. Thus, a user who replaces the output head can k...

📖 Read original article


162. LM-X: Explainable Action Modeling with Progress, Event, and Uncertainty Prediction for Generalist Robot Manipulation ​

Author: Jin Lou, Jingxuan Zhu, Andong Chen, Xupeng Wang, Yuan Xu, Yuexuan Li, Xingdong Zhu, Zhijie Zhu, Yingwei Ji, Wenpeng Nie, Jingyi Li, Liangliang Chen, Jinyan Liu, Zhiqi Song, Jidong Zhang, Hongming Li, Yuchen Zhu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.25757v1 Announce Type: cross Abstract: Generalist vision--language--action (VLA) policies learn long-horizon behavior mainly through short-horizon action prediction and reveal little beyond sampled commands. This creates two coupled bottlenecks: a single action target must implicitly abso...

📖 Read original article


163. Large Language Model Few-Shot Prompting with Dilemma Training Outperforms Human Surrogates in Predicting Patient Preferences ​

Author: Natasha Ureyang, Sebastian Porsdam Mann, Yuxin Liu, Zuriel Hassirim, Melanie Almonte, Wenhao Chen, Joyce Ng, Thant Nay Lin, Aung Thiha, Gerald CH Koh, Brian David Earp, Pin Sym Foong
Published: 8/27/2026, 4:00:00 AM
Categories: cs.HC, cs.LG

arXiv:2608.25771v1 Announce Type: cross Abstract: In serious illness, human surrogates often struggle to accurately predict patient preferences (68% accuracy), causing decision conflict. Personalized Patient Preference Predictor (P4) agents offer a potential solution, but prior prototypes treat valu...

📖 Read original article


164. TacForcing: Streaming Action Generation with Execution-Time Tactile Feedback ​

Author: Jianbo Zhou, Boyuan Zhao, Yuzheng Zhang, Yiyang Chen, Wenxin Chen, Qiuyue Li, Xiangyang Gu, Yuhan Cao, Xiao Xia, Yanzhe Hu, Zhijie Deng
Published: 8/27/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.25798v1 Announce Type: cross Abstract: Contact-rich manipulation requires adapting to contact states that can evolve substantially within an action horizon. However, chunk-based vision-language-action models predict complete action chunks from observations collected before execution, leav...

📖 Read original article


165. Unfolding Scientific Papers into Multi-Turn Generation Trajectories for Continued Pre-Training ​

Author: Qiankai Xu, Qiguang Chen, Zixin Su, Wenhao Huang, Yue Gao, Jiaheng Liu, Ge Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.25826v1 Announce Type: cross Abstract: A recent line of synthetic-data work reconstructs the thinking behind existing text rather than rewriting the text itself, but it operates on short web passages, recovers only local thoughts, and leaves the structure of whole documents untouched. Sci...

📖 Read original article


166. FlowMoDL: Model-Based Deep Learning with Conjugate-Gradient Data Consistency for Highly Accelerated 4D Flow MRI Reconstruction ​

Author: Tristan Gottwald, Michelle Bruch, Mubashir-Ul Hassan, Fatma Alickovic, Milan Kloiber, Daniel Tenbrinck, Torsten Panholzer, Melanie Schaller, Jana Hutter
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.25828v1 Announce Type: cross Abstract: We present FlowMoDL, an unrolled neural network for highly accelerated 4D flow MRI reconstruction that directly optimizes for both anatomical magnitude and phase-derived velocity accuracy. Building on the MoDL framework, FlowMoDL alternates a learned...

📖 Read original article


167. Skill Issue: Are Skills Language-Invariant in LLMs? ​

Author: Bobby Cheng, Adam Gaber, Zhengyuan Liu, Catherine Arnett, Omer Goldman, Cheston Tan, Leshem Choshen
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.GT, cs.LG

arXiv:2608.25832v1 Announce Type: cross Abstract: Large language models access knowledge inconsistently across languages, but to what extent do they differ in their skill sets when interacting with different languages? This work quantifies cross-lingual skill inconsistency orthogonally from knowledg...

📖 Read original article


168. Why ML-based cough models do not generalize: a systematic cross-dataset evaluation for tuberculosis screening ​

Author: Wensi Zhang, Tomas Teijeiro, J'er^ome Thevenot, David Atienza
Published: 8/27/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.LG, cs.SD, eess.SP

arXiv:2608.25846v1 Announce Type: cross Abstract: Cough acoustics are promising for non-invasive tuberculosis (TB) screening, yet whether machine learning (ML) models capture disease-related acoustics or artifacts of data collection remains unresolved. We evaluated the cross-dataset generalizability...

📖 Read original article


169. Key Point Analysis Needs Structure Recovery: Task Definition, Dataset Diagnosis, and a Structure-Aware Benchmark ​

Author: Zhiqiang Shi, Oana Cocarascu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.25854v1 Announce Type: cross Abstract: Key Point Analysis (KPA) aims to identify a concise set of key points that summarize a collection of arguments together with their prevalence. We argue that KPA is fundamentally a structured prediction problem that requires recovering semantic groupi...

📖 Read original article


170. Precipitation Downscaling Using Foundation Model-Conditioned Diffusion ​

Author: Victor Nascimento Ribeiro, Jorge Guevara, Jorge Sebastian Moraga, Chris Lucas, Natalie Lord, Andrew Taylor, Edward Lockhart, Will Trojak, Johannes Schmude, Anne Jones
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, physics.ao-ph

arXiv:2608.25858v1 Announce Type: cross Abstract: High-resolution precipitation fields are essential for hydrological impact assessment, yet global climate model outputs are too coarse and biased for direct use. AI-based statistical downscaling with diffusion models offers a promising approach, but ...

📖 Read original article


171. Efficient Estimation of High Information Projections using Nearest Neighbours ​

Author: David P. Hofmeyr
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.25887v1 Announce Type: cross Abstract: An intuitive method for dimensionality reduction is proposed, which is highly effective for finding interesting projections of multivariate data. Following similar intuitive motivation to a number of existing techniques, the proposed method is based ...

📖 Read original article


172. Scalable Multi-GPU Simulation of 3D Multicellular Growth with RNN-Based Workload Balancing ​

Author: Matvey Moisseyev, Huijing Du, Dandan Zheng, Chi Zhang, Hongfeng Yu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.DC, cs.CE, cs.LG, math.OC

arXiv:2608.25890v1 Announce Type: cross Abstract: Detailed multicellular growth simulations based on subcellular element models (SEMs) can capture complex tissue development, but their element-level interactions impose substantial computational cost. This work presents a scalable multi-GPU framework...

📖 Read original article


173. MetaSieve: Faster Relational Deep Learning through SQL-Based Metapath Selection ​

Author: Fahim Shahriar Khan, Ashraf Aboulnaga
Published: 8/27/2026, 4:00:00 AM
Categories: cs.DB, cs.LG

arXiv:2608.25903v1 Announce Type: cross Abstract: Relational Deep Learning (RDL) is an effective approach to machine learning over multi-table relational databases. In RDL, a database is modeled as a graph in which each row is a node and each foreign-key relation is an edge, and a graph neural netwo...

📖 Read original article


174. SAMpLE: A SystemC-AMS Machine LEarning-based Framework for Virtual Prototyping ​

Author: Andrei Mihai Albu, Sara Vinco
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.25910v1 Announce Type: cross Abstract: Machine Learning (ML) is increasingly used in virtual prototypes of embedded systems to model behaviors that are difficult to capture analytically. However, integrating ML models into virtual platform simulation is still typically done through ad hoc...

📖 Read original article


175. Controlling for Omitted Variable Bias in Deep Neural Networks ​

Author: Manuel Pfeuffer, Roshan Prakash Rane, Kerstin Ritter, Sonja Greven
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ME, cs.CV, cs.LG

arXiv:2608.25930v1 Announce Type: cross Abstract: Control variables are widely used in statistical modelling to account for omitted variable bias of known confounders. However, they have largely been underexplored in deep learning. This is surprising, given that deep learning models encode image-inf...

📖 Read original article


176. Continually learning neural-operator surrogate for three-dimensional airborne electromagnetic Bayesian inversion ​

Author: Jaehong Chung, Andrew Lockwood, Jef Caers
Published: 8/27/2026, 4:00:00 AM
Categories: physics.geo-ph, cs.LG

arXiv:2608.25932v1 Announce Type: cross Abstract: Three-dimensional probabilistic inversion of time-domain airborne electromagnetic (AEM) data is limited by the cost of the forward solve. Even though one simulation takes only tens of seconds, a Bayesian inversion of a survey of millions of soundings...

📖 Read original article


177. How Robust Are Automated Fact-Checking Systems? A Cross-Benchmark Evaluation ​

Author: Aida Usmanova, Zangir Iklassov, Markus Leippold, Ricardo Usbeck
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.25934v1 Announce Type: cross Abstract: Automated fact-checking (AFC) systems retrieve evidence and predict claim veracity, yet evaluations omit simple baselines, systems are developed for a single benchmark and cannot be trusted to generalise across domains. No prior work cross-evaluates ...

📖 Read original article


178. SciMIF: Understanding Multimodal Instruction Following in Scientific Domains ​

Author: Ye Shen, Yuting Zheng, Dun Pei, Zijian Chen, Wenlong Zhang, Qi Jia, Guangtao Zhai
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.25973v1 Announce Type: cross Abstract: Understanding instruction-following capabilities in scientific domains is essential for effectively leveraging Multimodal Large Language Models (MLLMs) to advance the development of scientific fields. In this work, we introduce SciMIF, a novel benchm...

📖 Read original article


179. Lost but not erased: Finding traces of a forgotten language in neural speech models ​

Author: Peter Plantinga, Charlotte Moore, Peter W. Donhauser, Krista Byers-Heinlein, Denise Klein
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.25976v1 Announce Type: cross Abstract: International adoptees retain phonological traces of a birth language they can no longer speak or comprehend, a persistence typically attributed to a biologically-timed critical period. We asked whether it could instead reflect the ordinary dynamics ...

📖 Read original article


180. FRAME: separating sampling variation from representational cause in medical imaging fairness ​

Author: Mahshad Lotfinia, Daniel Truhn, Andreas Maier, Soroosh Tayebi Arasteh
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.25981v1 Announce Type: cross Abstract: Subgroup performance differences are the standard evidence for fairness bias in medical imaging, and the usual response removes the demographic information that a model encodes. Here we introduce Fair-model Reference And Mechanism Evaluation (FRAME),...

📖 Read original article


181. CardioFusion-AI: Robust ECG--PPG Fusion for Multimodal Physiological Monitoring Under Signal Degradation ​

Author: Navaneetha Krishnan Kamalakannan, Janakiraman Kamalakannan
Published: 8/27/2026, 4:00:00 AM
Categories: eess.SP, cs.LG, q-bio.QM

arXiv:2608.26000v1 Announce Type: cross Abstract: Wearable electrocardiogram (ECG) and photoplethysmogram (PPG) sensors are complementary but individually fragile: motion artifact, poor contact, and sensor dropout can degrade one or both signals. Fusion strategies that assume both modalities are equ...

📖 Read original article


182. Imitation Learning for Connection-Tableau Construction ​

Author: Fredrik R{\o}mming, Mantas Bak\v{s}ys, Martin S. Fixman, Sean B. Holden
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO

arXiv:2608.26009v1 Announce Type: cross Abstract: An automated theorem prover builds a proof step by step, choosing at each point what to add and what to remove. We cast this construction as a policy acting in a transition system induced by a formal calculus, which fixes which steps are sound: for c...

📖 Read original article


183. $R^3$: Training Robots to Reason in Natural Language via Reinforcement Learning ​

Author: Lehong Wu, Yuxiao Qu, Zheyuan Hu, Ivan Zhang, Limin Wei, Zackory Erickson, Aviral Kumar
Published: 8/27/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CL, cs.LG

arXiv:2608.26053v1 Announce Type: cross Abstract: Reasoning in language allows foundation models to spend more test-time compute on hard problems, such as those requiring decomposition, constraint tracking, and prediction of future consequences. Whether this mechanism can improve robotic manipulatio...

📖 Read original article


184. Prefix Sliding for efficient test-time scaling ​

Author: Niklas Muennighoff, Zhengyang Wang, Zeyi Chen, Weijia Shi, Binyuan Hui, John Yang, Dapeng Jiang, Mika Senghaas, Fares Obeid, Johannes Hagemann, Sami Jaghouar, Ludwig Schmidt, Percy Liang, Jason Wei, Andrew Y. Ng, Luke Zettlemoyer, Yejin Choi, Mike Lewis
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.26070v1 Announce Type: cross Abstract: Test-time scaling uses extra test-time compute to improve performance, such as letting language models reason longer when solving a problem. As models keep the entire reasoning trace in memory via full attention, hard tasks that need long thinking ca...

📖 Read original article


185. Planetary Prediction Engine: Autonomous Geospatial Prediction via Intelligent Data Selection and Foundation Model Embeddings ​

Author: Evelyn Ma, Rama Kumar Pasumarthi, Kishwar Shafin, Mandar Sharma, Mimi Sun, Hamed Sadeghi, Dav M. Ebengo, Mbulayi Onesime, Rouslan Solomakhin, John Wamburu, William Ogallo, Aisha Walcott-Bryant, Sanxing Chen, Arbaaz Muslim, Yael Mayer, Ronald Ho, Roy Lee, Ruth Alcantara, Abdoulaye Diack, Monica Bharel, Lambert Rosique, Jeremy Amez-Droz, Christopher Haire, James Manyika, Yossi Matias, Niv Efron, Gautam Prasad, Shravya Shetty
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.26088v1 Announce Type: cross Abstract: Addressing critical global challenges, from food security and disaster risk to disease outbreaks and socio-economic vulnerability, demands high-fidelity geospatial modeling. However, building predictive planetary models remains bottlenecked by a frag...

📖 Read original article


186. Finding and using interpretable latents in a neutrino foundation model with sparse autoencoders ​

Author: Rapha"el Bonnet-Guerrini, Johann Ioannou-Nikolaides, Inar Timiryasov, Vincenzo Piuri
Published: 8/27/2026, 4:00:00 AM
Categories: astro-ph.HE, cs.AI, cs.LG, hep-ex

arXiv:2608.26090v1 Announce Type: cross Abstract: We present a first application of sparse-autoencoder-based mechanistic interpretability to particle physics. Studying a neutrino foundation model pretrained on IceCube data and fine-tuned for direction reconstruction, we identify a validated atlas of...

📖 Read original article


187. MyoMechanix: Biomechanically-Grounded Compositional Skilled Activity Understanding and Coaching ​

Author: Hao Yin, Paritosh Parmar, Lijun Gu, Lin Xu, Tianxiao Guo, Xiujin Liu, Tianyou Zheng, Yang Zhang, Weiwei Fu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.ET, cs.HC, cs.LG

arXiv:2608.26094v1 Announce Type: cross Abstract: Existing action quality assessment (AQA) datasets and methods rely primarily on visual inputs such as RGB and pose, overlooking physiological dynamics such as muscle mechanics and often modeling actions as monolithic patterns. These limitations hinde...

📖 Read original article


188. VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning ​

Author: Junxiang Xu, Ruisi Wang, Fanyi Pu, Maijunxian Wang, Ran Ji, Tongxi Zhou, Chenyang Gu, Jing Zuo, Hongcan Xiao, Yimeng Geng, Wanqi Yin, Wei Chen, Oscar Qian, Zhengan Yan, Ziqi Huang, Haiwen Diao, Liang Pan, Bo Li, Xiangyu Fan, Dezhi Luo, Fengyuan Yu, Zehong Zhao, Qingying Gao, Tinghui Zhu, Yilan Zhang, Jingqi Tong, Pinyuan Feng, Zhengze Jiang, Letian Wang, Ziyu Guo, Renrui Zhang, Jieneng Chen, Sonia Joseph, Constantin Venhoff, Saman Motamed, Mengyue Yang, Chandra Sripada, Alan Yuille, Philip Torr, Lvmin Zhang, Vikash Kumar, Daniel Khashabi, Nikolaus Kriegeskorte, Rapha"el Milli`ere, Vincent C. M"uller, Anyi Rao, Quan Wang, Ziwei Liu, Dahua Lin, Lei Yang, Hokin Deng, Zhongang Cai
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MM, cs.RO

arXiv:2608.26105v1 Announce Type: cross Abstract: Native visual reasoning treats visual generation as the medium of reasoning itself: visual states (i.e. images and videos) are not merely inputs to be understood or outputs to be rendered, but first-class substrates for problem solving beyond languag...

📖 Read original article


189. Theoretically Principled Federated Learning for Balancing Privacy and Utility ​

Author: Xiaojin Zhang, Wenjie Li, Yiming Li, Wei Chen, Shutao Xia, Qiang Yang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2305.15148v3 Announce Type: replace Abstract: We propose a general learning framework for the protection mechanisms that protects privacy via distorting model parameters, which facilitates the trade-off between privacy and utility. The algorithm is applicable to arbitrary privacy measurements ...

📖 Read original article


190. Noise Contrastive Estimation-based Matching Framework for Low-Resource Security Attack Pattern Recognition ​

Author: Tu Nguyen, Nedim \v{S}rndi'c, Alexander Neth
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CR

arXiv:2401.10337v5 Announce Type: replace Abstract: Tactics, Techniques and Procedures (TTPs) represent sophisticated attack patterns in the cybersecurity domain, described encyclopedically in textual knowledge bases. Identifying TTPs in cybersecurity writing, often called TTP mapping, is an impor...

📖 Read original article


191. Differentiated Aggregation to Improve Generalization in Federated Learning ​

Author: Peyman Gholami, Hulya Seferoglu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2404.11754v4 Announce Type: replace Abstract: This paper focuses on reducing the communication cost of federated learning by exploring generalization bounds and representation learning. We first characterize a tighter generalization bound for one-round federated learning based on local clients...

📖 Read original article


192. Provable Privacy Attacks on Trained Shallow Neural Networks ​

Author: Guy Smorodinsky, Gal Vardi, Itay Safran
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2410.07632v3 Announce Type: replace Abstract: We study what provable privacy attacks can be shown for trained 2-layer ReLU neural networks, focusing on two types of attacks: membership inference and data reconstruction. We prove that theoretical results on the implicit bias of 2-layer neural n...

📖 Read original article


193. Optimal Time Complexity Algorithms for Computing General Random Walk Graph Kernels on Sparse Graphs ​

Author: Krzysztof Choromanski, Isaac Reid, Arijit Sehanobish, Avinava Dubey
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2410.10368v3 Announce Type: replace Abstract: We present the first linear time complexity randomized algorithms for unbiased approximation of the celebrated family of general random walk kernels (RWKs) for sparse graphs. This includes both labelled and unlabelled instances. The previous fastes...

📖 Read original article


194. A General-Purpose Framework for Chemical Reaction Representation with Atomic Correspondence and Flexible Condition Adaptation ​

Author: Kaipeng Zeng, Xianbin Liu, Yu Zhang, Xiaokang Yang, Yaohui Jin, Yanyan Xu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2411.17629v3 Announce Type: replace Abstract: Motivation: Organic synthesis is fundamental to the chemical industry, particularly in domains such as pharmaceutical development. While artificial intelligence offers powerful tools for modeling chemical reactions, current approaches are primarily...

📖 Read original article


195. DeltaGNN: Graph Neural Network with Information Flow Control ​

Author: Kevin Mancini, Islem Rekik
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2501.06002v3 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are popular deep learning models designed to process graph-structured data through recursive neighborhood aggregations in the message passing process. When applied to semi-supervised node classification, the message-pas...

📖 Read original article


196. BRIDLE: Generalized Self-supervised Learning with Quantization ​

Author: Hoang M. Nguyen, Satya N. Shukla, Qiang Zhang, Hanchao Yu, Sreya D. Roy, Dipesh Tamboli, Taipeng Tian, Lingjiong Zhu, Yuchen Liu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2502.02118v2 Announce Type: replace Abstract: Self-supervised learning has been a powerful approach for learning meaningful representations from unlabeled data across various domains, reducing the reliance on large labeled datasets. Inspired by BERT's success in capturing deep bidirectional co...

📖 Read original article


197. BAGEL: Adversarially Constrained Online Convex Optimization under Separation Oracle Access ​

Author: Yiyang Lu, Mohammad Pedramfar, Mengbo Wang, Vaneet Aggarwal
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC

arXiv:2502.16744v3 Announce Type: replace Abstract: In adversarial Constrained Online Convex Optimization (COCO), a learner selects actions from a fixed convex set while seeking both low regret and low cumulative constraint violation (CCV) under time-varying constraints. We ask what performance is a...

📖 Read original article


198. Emergent Abilities in Large Language Models: A Survey ​

Author: Leonardo Berti, Flavio Giorgi, Gjergji Kasneci
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2503.05788v3 Announce Type: replace Abstract: Large Language Models (LLMs) are leading a new technological revolution as one of the most promising research streams toward artificial general intelligence. The scaling of these models, accomplished by increasing the number of parameters and the m...

📖 Read original article


199. Deep greedy unfolding: Sorting out argsorting in greedy sparse recovery algorithms ​

Author: Sina Mohammad-Taheri, Matthew J. Colbrook, Simone Brugiapaglia
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, cs.NE, math.NA

arXiv:2505.15661v2 Announce Type: replace Abstract: Gradient-based learning imposes (deep) neural networks to be differentiable at all steps. This includes model-based architectures constructed by unrolling iterations of an iterative algorithm onto layers of a neural network, known as algorithm unro...

📖 Read original article


200. Scorpio: Serving Right Requests at the Right Time for Heterogeneous SLOs in LLM Inference ​

Author: Yinghao Tang, Tingfeng Lan, Bo Pan, Xiuqi Huang, Hui Lu, Wei Chen
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2505.23022v2 Announce Type: replace Abstract: Large Language Model (LLM) serving increasingly underpins online Web services such as conversational agents, Web search, and programming assistants, where requests carry heterogeneous Service Level Objectives (SLOs) such as Time to First Token (TTF...

📖 Read original article


201. Sample Margin-Aware Recalibration of Temperature Scaling ​

Author: Haolan Guo, Linwei Tao, Haoyang Luo, Minjing Dong, Chang Xu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2506.23492v2 Announce Type: replace Abstract: Recent advances in deep learning have significantly improved predictive accuracy. However, modern neural networks remain systematically overconfident, posing risks for deployment in safety-critical scenarios. Current post-hoc calibration methods fa...

📖 Read original article


202. Learning to summarize user information for personalized reinforcement learning from human feedback ​

Author: Hyunji Nam, Yanming Wan, Mickel Liu, Peter Ahnn, Jianxun Lian, Natasha Jaques
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2507.13579v4 Announce Type: replace Abstract: As everyday use cases of large language model (LLM) AI assistants have expanded, it is becoming increasingly important to personalize responses to align to different users' preferences and goals. While reinforcement learning from human feedback (RL...

📖 Read original article


203. AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning ​

Author: Can Jin, Yang Zhou, Qixin Zhang, Hongwu Peng, Di Zhang, Zihan Dong, Marco Pavone, Ligong Han, Zhang-Wei Hong, Tong Che, Dimitris N. Metaxas
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2508.14313v5 Announce Type: replace Abstract: Test-time scaling strategies for Large Language Models predominantly rely on either reinforcement learning with sparse outcome rewards or search-based methods guided by static Process Reward Models. However, outcome-based RL often suffers from trai...

📖 Read original article


204. Ban&Pick: Enhancing Performance and Efficiency of MoE-LLMs via Smarter Routing ​

Author: Yuanteng Chen, Peisong Wang, Yuantian Shao, Nanxin Zeng, Chang Xu, Jian Cheng
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2509.06346v3 Announce Type: replace Abstract: Sparse Mixture-of-Experts (MoE) has become a key architecture for scaling large language models (LLMs) efficiently. Recent fine-grained MoE designs introduce hundreds of experts per layer, with multiple experts activated per token, enabling stronge...

📖 Read original article


205. CountTRuCoLa: Rule Learning for Interpretable Temporal Knowledge Graph Forecasting ​

Author: Julia Gastinger, Christian Meilicke, Heiner Stuckenschmidt
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.09474v4 Announce Type: replace Abstract: We address the task of temporal knowledge graph forecasting with an inherently interpretable method based on symbolic rules. Motivated by recent work proposing a strong baseline based on recurrent facts, our approach learns four simple rule types, ...

📖 Read original article


206. Rotary Position Encodings for Graphs ​

Author: Isaac Reid, Arijit Sehanobish, Cederik H"ofs, Bruno Mlodozeniec, Leonhard Vulpius, Federico Barbero, Adrian Weller, Krzysztof Choromanski, Richard E. Turner, Petar Veli\v{c}kovi'c
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2509.22259v5 Announce Type: replace Abstract: We study the extent to which rotary position encodings (RoPE), a recent transformer position encoding algorithm broadly adopted in large language models (LLMs) and vision transformers (ViTs), can be applied to graph-structured data. We find that ro...

📖 Read original article


207. ONNX-Net: Towards Universal Representations and Instant Performance Prediction for Neural Architectures ​

Author: Shiwen Qin, Alexander Auras, Shay B. Cohen, Elliot J. Crowley, Michael Moeller, Linus Ericsson, Jovita Lukasik
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2510.04938v2 Announce Type: replace Abstract: Neural architecture search (NAS) automates the design process of high-performing architectures, but remains bottlenecked by expensive performance evaluation. Most existing studies that achieve faster evaluation are mostly tied to cell-based search ...

📖 Read original article


208. AlgoTrace: Algorithmic Primitives and Compositional Geometry of Reasoning in Language Models ​

Author: Samuel Lippl, Thomas McGee, Kimberly Lopez, Ziwen Pan, Pierce Zhang, Salma Ziadi, Oliver Eberle, Ida Momennejad
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.15987v3 Announce Type: replace Abstract: How do inference time and latent computations enable large language models (LLMs) to solve multi-step reasoning problems? We introduce AlgoTrace, a framework for tracing and steering algorithmic operations in the model latent space for multi-step r...

📖 Read original article


209. Epistemic Memory: A Validity Layer for Self-Maintaining Intelligent Systems ​

Author: Pin-Han Ho, Limei Peng, Yiming Miao, Yan Jiao
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.16899v2 Announce Type: replace Abstract: AI memory mechanisms primarily focus on preserving information content, often neglecting the validity conditions under which knowledge remains applicable, leading to semantic coordinate drift when agents move, change sensors, or encounter novel env...

📖 Read original article


210. Cluster-Dags as Powerful Background Knowledge For Causal Discovery ​

Author: Jan Marco Ruiz de Vargas, Kirtan Padh, Niki Kilbertus
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2512.10032v3 Announce Type: replace Abstract: Finding cause-effect relationships is of key importance in science. Causal discovery aims to recover a graph from data that succinctly describes these cause-effect relationships. However, current methods face several challenges, especially when dea...

📖 Read original article


211. A Comedy of Estimators: On KL Regularization in RL Training of LLMs ​

Author: Vedant Shah, Johan Obando-Ceron, Vineet Jain, Brian Bartoldson, Bhavya Kailkhura, Sarthak Mittal, Glen Berseth, Pablo Samuel Castro, Yoshua Bengio, Esmeralda S. Whitammer, Moksh Jain, Siddarth Venkatraman, Aaron Courville
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2512.21852v4 Announce Type: replace Abstract: The reasoning performance of large language models (LLMs) can be substantially improved by training them with reinforcement learning (RL). The RL objective for LLM training involves a regularization term, which is the reverse Kullback-Leibler (KL) ...

📖 Read original article


212. Predicting Time Pressure of Powered Two-Wheeler Riders for Proactive Safety Interventions ​

Author: Sumit S. Shevtekar, Chandresh K. Maurya, Gourab Sil
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.HC

arXiv:2601.03173v4 Announce Type: replace Abstract: Time pressure critically influences risky maneuvers and crash proneness among powered two-wheeler riders, yet its prediction remains underexplored in intelligent transportation systems. To address this gap, we propose MotoTimePressure (MTPS), a dee...

📖 Read original article


213. StablePDENet: Enhancing Neural Operator Stability through Physics-Informed Residual-Sensitivity Regularization ​

Author: Chutian Huang, Chang Ma, Kaibo Wang, Yang Xiang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.06472v2 Announce Type: replace Abstract: Learning solution operators for differential equations with neural networks has shown great potential in scientific computing, but ensuring their stability under input perturbations remains a critical challenge. We introduce the StablePDENet, a phy...

📖 Read original article


214. GRIP: Algorithm-Agnostic Machine Unlearning for Mixture-of-Experts via Geometric Router Constraints ​

Author: Andy Zhu, Rongzhe Wei, Yupu Gu, Pan Li
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.16905v3 Announce Type: replace Abstract: Machine unlearning in Mixture-of-Experts (MoE) large language models presents a critical yet under-explored challenge. Current unlearning methods applied to MoE architectures often exploit dynamic routing as an optimization shortcut: rather than ge...

📖 Read original article


215. Loss Landscape Geometry of Partial Differential Equation Emulators: Or, Symmetry Learning via Gradient Alignment ​

Author: James Amarel, Robyn Miller, Nicolas Hengartner, Benjamin Migliori, Emily Casleton, Alexei Skurikhin, Earl Lawrence, Gerd J. Kunde
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph

arXiv:2601.20172v2 Announce Type: replace Abstract: We study how neural emulators of partial differential equation solution operators internalize physical symmetries by introducing an influence-based diagnostic that measures the propagation of parameter updates between symmetry-related states, defin...

📖 Read original article


216. SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning ​

Author: Zhen-Hao Xie, Jun-Tao Tang, Yu-Cheng Shi, Han-Jia Ye, De-Chuan Zhan, Da-Wei Zhou
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.01990v3 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance through instruction tuning, but real-world deployment requires them to continually expand their capabilities, making Multimodal Continual Instruction Tuning (MCIT) essential. Recen...

📖 Read original article


217. Maximum-Volume Nonnegative Matrix Factorization ​

Author: Olivier Vu Thanh, Nicolas Gillis
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, eess.SP, math.NA, stat.ML

arXiv:2602.04795v3 Announce Type: replace Abstract: Nonnegative matrix factorization (NMF) is a popular data embedding technique. Given a nonnegative data matrix $X$, it aims at finding two lower dimensional matrices, $W$ and $H$, such that $X\approx WH$, where the factors $W$ and $H$ are constraine...

📖 Read original article


218. Spatio-temporal dual-stage hypergraph MARL for human-centric multimodal corridor traffic signal control ​

Author: Xiaocai Zhang, Neema Nassir, Milad Haghani
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2602.17068v2 Announce Type: replace Abstract: Human-centric traffic signal control in corridor networks must increasingly account for multimodal travelers, particularly high-occupancy public transportation, rather than focusing solely on vehicle-centric performance. This paper proposes STDSH-M...

📖 Read original article


219. Phase-Consistent Magnetic Spectral Learning for Multi-View Clustering ​

Author: Mingdong Lu, Zhikui Chen, Meng Liu, Shubin Ma, Zhengyang Tang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2602.18728v2 Announce Type: replace Abstract: Unsupervised multi-view clustering (MVC) aims to partition data into meaningful groups by leveraging complementary information from multiple views without labels, yet a central challenge is to obtain a reliable shared structural signal to guide rep...

📖 Read original article


220. Multi-Turn Reasoning LLMs for Task Offloading in Mobile Edge Computing ​

Author: Ning Yang, Chuangxin Cheng, Haijun Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.07148v2 Announce Type: replace Abstract: Emerging computation-intensive applications impose stringent latency requirements on resource-constrained mobile devices. Mobile Edge Computing (MEC) addresses this challenge through task offloading. However, designing effective policies remains di...

📖 Read original article


221. Continuous Adversarial Flow Models ​

Author: Shanchuan Lin, Ceyuan Yang, Zhijie Lin, Hao Chen, Haoqi Fan
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2604.11521v2 Announce Type: replace Abstract: We propose continuous adversarial flow models, a type of continuous-time flow model trained with an adversarial objective. Unlike flow matching, which uses a fixed mean-squared-error criterion, our approach introduces a learned discriminator to gui...

📖 Read original article


222. A Layer-wise Analysis of Supervised Fine-Tuning ​

Author: Qinghua Zhao, Xueling Gong, Xinyu Chen, Zhongfeng Kang, Xinlu Li
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.11838v2 Announce Type: replace Abstract: While critical for alignment, Supervised Fine-Tuning (SFT) incurs the risk of catastrophic forgetting, yet the layer-wise emergence of instruction-following capabilities remains elusive. We investigate this mechanism via a comprehensive analysis ut...

📖 Read original article


223. Loop Corrections in Random Feature Models: Training Error and Generalization Gap ​

Author: Taeyoung Kim
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2604.12827v4 Announce Type: replace Abstract: We study fixed-design random feature ridge regression beyond the mean-kernel approximation. The expectation is taken over the frozen-feature ensemble, conditional on the training sample. Because the predictor is a nonlinear function of the empirica...

📖 Read original article


224. StoSignSGD: Unbiased Structural Stochasticity Fixes SignSGD for Training Large Language Models ​

Author: Dingzhi Yu, Rui Pan, Yuxing Liu, Difan Zou, Tong Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC

arXiv:2604.15416v2 Announce Type: replace Abstract: Sign-based optimization algorithms, such as SignSGD, have garnered attention for their performance in distributed learning and training large foundation models. Despite their empirical superiority, SignSGD is known to diverge on non-smooth objectiv...

📖 Read original article


225. Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models ​

Author: Jeongjae Lee, Jinho Chang, Jeongsol Kim, Jong Chul Ye
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2604.17415v4 Announce Type: replace Abstract: Reward-based fine-tuning steers a pretrained diffusion or flow-based generative model toward higher-reward samples while remaining close to the pretrained model. Although existing methods are derived from different perspectives, we show that many c...

📖 Read original article


226. JEPAMatch: Geometric Representation Shaping for Semi-Supervised Learning ​

Author: Ali Aghababaei-Harandi, Aude Sportisse, Massih-Reza Amini
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.21046v3 Announce Type: replace Abstract: Semi-supervised learning has emerged as a powerful paradigm for leveraging large amounts of unlabeled data to improve the performance of machine learning models when labeled data are scarce. Among existing approaches, methods derived from FixMatch ...

📖 Read original article


227. Towards Robust and Scalable Density-based Clustering via Graph Propagation ​

Author: Yingtao Zheng, Hugo Phibbs, Ninh Pham
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.00390v2 Announce Type: replace Abstract: We present \textit{CluProp}, a novel framework that reimagines varied-density clustering in high-dimensional spaces as a label propagation process over neighborhood graphs. Our approach formally bridges the gap between density-based clustering and ...

📖 Read original article


228. Cubit: Token Mixer with Kernel Ridge Regression ​

Author: Chuanyang Zheng, Jiankai Sun, Yihang Gao, Yuehao Wang, Liangchen Tan, Mac Schwager, Anderson Schneider, Yuriy Nevmyvaka, Xiaodong Liu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2605.06501v3 Announce Type: replace Abstract: Since its introduction in 2017, the Transformer has become one of the most widely adopted architectures in modern deep learning. Despite extensive efforts to improve positional encoding, attention mechanisms, and feed-forward networks, the core tok...

📖 Read original article


229. Amplifying, Not Learning: The Price of Out-of-Distribution Generalization in AI-Text Detection ​

Author: Alexander Smirnov
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2605.21653v2 Announce Type: replace Abstract: AI-text detectors gate decisions in education, hiring, and publishing, yet they flag the most fluent, formal human writing as machine-generated: they rate the median formal-native human essay as 99.5% likely AI while clearing genuine high-temperatu...

📖 Read original article


230. GlucoFM: A Dual-Stream Foundation Model for Continuous Glucose Monitoring ​

Author: Zechen Li, Keerthana Natarajan, Weizhi Zhang, Menglian Zhou, Simon A. Lee, Yuwei Zhang, Maxwell A. Xu, Zeinab Esmaeilpour, Flora D. Salim, Mark Malhotra, Lindsey Sunden, Shwetak Patel, Yuzhe Yang, Ahmed A. Metwally
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.30865v2 Announce Type: replace Abstract: Continuous glucose monitoring (CGM) provides a dense view of daily metabolic physiology, yet existing generic time-series and CGM-specific foundation models often encode glucose traces as entangled single-stream sequences, leaving their multiscale ...

📖 Read original article


231. Mitigating False Credit Propagation: Probabilistic Graphical Reward Aggregation for Rubric-Based Reinforcement Learning ​

Author: Can Lv, Mingju Chen, Heng Chang, Shiji Zhou
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.03361v2 Announce Type: replace Abstract: Rubric-based rewards are increasingly used for open-ended language model post-training, but criterion-level scores are often aggregated as independent utilities. This flat scalarization ignores rubric-specified prerequisite and activation relations...

📖 Read original article


232. MODE: Modality-Decomposed Expert-Level Mixed-Precision Quantization for MoE Multimodal LLMs ​

Author: Yuanteng Chen, Nanxin Zeng, Peisong Wang, Zhilei Liu, Yuantian Shao, Shiqiang Lang, Tao Liu, Chuangyi Li, Qinghao Hu, Gang Li, Jing Liu, Jian Cheng
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.17118v2 Announce Type: replace Abstract: Mixture-of-Experts Multimodal Large Language Models (MoE-MLLMs) offer remarkable performance but incur prohibitive GPU memory costs, making compression essential. Among PTQ methods, expert-level mixed-precision quantization has proven effective for...

📖 Read original article


233. Emyx: Fast and efficient all-atom protein generation ​

Author: Nicholas J. Williams, Ward Haddadin, Matteo P. Ferla, Constantin Schneider, Nicholas B. Woodall, Ruby Sedgwick, Christian D. Madsen, Andrew L. Hopkins, Douglas E. V. Pires, Edward O. Pyzer-Knapp
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.19377v2 Announce Type: replace Abstract: Computational enzyme design requires generating proteins that scaffold catalytic residues and ligands, a task that demands both geometric accuracy and structural diversity from the underlying generative model. Current all-atom generators inherit ex...

📖 Read original article


234. Machine-learnable Sets ​

Author: Veit Elser, Manish Krishan Lal
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.28947v2 Announce Type: replace Abstract: In this study we present a formal definition of large discrete sets having, informally, three properties: their elements are easily recognized, easily generated, and the latter tasks are easily learned from examples. The formalism is specialized to...

📖 Read original article


235. Adaptive Bayes exactly tracks information over intrinsic time ​

Author: Akshay Balsubramani
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, math.ST, stat.ML, stat.TH

arXiv:2607.08789v2 Announce Type: replace Abstract: Bayesian and multiplicative-weights updates reweight experts, models, or actions from sequential feedback. We show that the regret of any such update obeys an exact information-accounting identity. On each round, the learner's excess loss to any ch...

📖 Read original article


236. Activation Steering Transfer to Agents: One Gain Ratio Does Not Identify Potency and Efficacy ​

Author: Lucas Pinto
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09156v2 Announce Type: replace Abstract: Additive activation steering is calibrated in single-turn chat and then deployed inside agent scaffolds. The quantity usually reported for that move is a gain: a ratio of steered effects, T = Delta_agent / Delta_chat. We sweep eight family x arm do...

📖 Read original article


237. Measuring the Dependency Gap: Diagnosing Inter-Column Fidelity in Tabular Generative Models ​

Author: Jie Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.21636v5 Announce Type: replace Abstract: Synthetic tabular data are valued for preserving inter-column dependency, yet each routine fidelity score is a single number that says neither where that dependency is lost nor why. We localize the deficit inside a single score. Equipping a classif...

📖 Read original article


238. Temporally Centered SIGReg Improves LeWorldModel Representations for Robot Policy Learning ​

Author: Chang Liu, Fei Suo, Yanzhou Jin, Zeyu Ping, Yusuke Iwasawa, Yutaka Matsuo, Yaonan Zhu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2607.26924v3 Announce Type: replace Abstract: Recent work on LeWorldModel (LeWM) has shown that the Sketched Isotropic Gaussian Regularizer (SIGReg) enables stable end-to-end world model learning from pixels by regularizing the latent representation toward an isotropic Gaussian. While effectiv...

📖 Read original article


239. Adaptivity via a Parallel Architecture for Stochastic Gradient Methods ​

Author: Bin Fu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, math.OC

arXiv:2607.28902v2 Announce Type: replace Abstract: Let $\mathrm{A}(x_0,y)$ be an algorithm with two inputs: an initial point $x_0$ and an integer parameter $y$, which specifies that $\mathrm{A}(.,.)$ executes at most $y$ iterations or steps. Given an integer $p\ge 1$, $p$ parallel processors search...

📖 Read original article


240. Riemannian Attention Mechanisms for Transformers: A Theoretical Framework and Architecture Design ​

Author: Sen Song
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.01283v2 Announce Type: replace Abstract: All Transformer-based large language models compute attention via the Euclidean inner product, an architectural choice that Dong et al. (2021) proved causes representational rank to decay doubly exponentially with depth in pure self-attention stack...

📖 Read original article


241. Contrastive Learning for Interpretable Anomaly Detection at Collider Experiments ​

Author: Haoyi Jia, Sagar Addepalli, Julia Gonski
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, hep-ex, hep-ph

arXiv:2608.13652v2 Announce Type: replace Abstract: Generic event-level anomaly detection for collider physics has two recurring problems: anomaly scores are hard to interpret, and they correlate strongly with energy scale and object multiplicity. We present Organized Representation via Contrastive ...

📖 Read original article


242. Mitigating Rubric Interference in LLM Judges via On-Policy Self-Distillation ​

Author: Dingyao Yu, Tong Zhang, Yutao Mou, Yunxiao Zhang, Wei Ye, Shikun Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14684v2 Announce Type: replace Abstract: LLM judges increasingly evaluate responses against fine-grained rubric checklists. When a sample requires multiple rubrics, current methods typically assess each in a separate inference call. Evaluating all rubrics in a single pass is a natural alt...

📖 Read original article


243. Towards a theory of inference-time alignment with unknown rewards ​

Author: Steve Hanneke, Hongao Wang, Mingyue Xu
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15402v3 Announce Type: replace Abstract: Generative model alignment has received broad interest, and significant progress has been made in supervised fine-tuning and inference-time computation. Yet, alignment has remained poorly understood from a statistical learning perspective. We formu...

📖 Read original article


244. SCALE: State-Calibrated Latent Embeddings for JEPA Planning in the Right Geometry ​

Author: Jiaming Hu, Yan Zheng, Tian Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16287v2 Announce Type: replace Abstract: Joint-embedding predictive world models plan by scoring predicted terminal embeddings against a goal embedding using a cost defined on the representation itself. Two prominent strategies for obtaining non-collapsed representations are to inherit a ...

📖 Read original article


245. MotoSafety: Edge-AI with Learned Temporal Importance for Two-Wheeler Collision Risk Assessment Under Time Pressure ​

Author: Sumit S. Shevtekar, Chandresh K. Maurya, Gourab Sil, Subasish Das
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC

arXiv:2608.17823v3 Announce Type: replace Abstract: Powered two-wheeler riders face critical safety challenges in low- and middle-income countries, yet limited studies exist on how cognitive stressors such as Time Pressure influence collision risk. We address this gap by introducing a comprehensive ...

📖 Read original article


246. Uncovering the Limits of Proof Sharing for Neural Networks ​

Author: Kanak Das, Shubham Ugare, Bor-Yuh Evan Chang, Sasa Misailovic, Gagandeep Singh, Manu Sridharan
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.19351v2 Announce Type: replace Abstract: Robustness verification of neural networks is increasingly important, due to their use in many critical domains. In certain scenarios, proof sharing has been shown to accelerate incomplete verification techniques by reusing intermediate-layer abstr...

📖 Read original article


247. Multi-Source Complex Network Reconstruction via Wasserstein Distributionally Robust Optimization and Algorithm Unrolling ​

Author: Chuansen Peng, Yifan Xia, Jinshan Zhong, Xiaojing Shen
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.19914v3 Announce Type: replace Abstract: Reconstructing complex network topologies from data is a fundamental challenge in cybernetics and graph signal processing, with applications in neuroscience, sensor, and social networks. In practice, target-domain samples are scarce while heterogen...

📖 Read original article


248. Scaling Muon for Diffusion Transformers ​

Author: Chenghao Li, Xiao Han, Xinxin Huang, Wei Liu, Boyang Li, Bing Xiao, Heran Zhang, Juanma Perez Rua, Ke Xu, Kangning Liu, Linjun Kuang, Na Li, Tan Wang, Tian Xie, Wei Peng, Yang Pei, Yifan Xu, Yuanhao Zhai, Yuwei Lin, Zhe Wang, Zihao He, Daniel Li, Junbiao Tang, Ziyang Jiang, Dake Chen
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.20818v3 Announce Type: replace Abstract: The matrix-aware optimizer Muon improves large model training by balancing updates across singular directions, yet its scaling behavior and end-to-end efficiency on large Diffusion Transformers (DiTs) remain unclear. We first establish Muon's scali...

📖 Read original article


249. Trojaning the Alignment: Stealthy Backdoor Attacks against Graph Foundation Models ​

Author: Minhua Lin, Zhicheng Gao, Yilong Wang, Hanqing Lu, Xiang Zhang, Suhang Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.20991v2 Announce Type: replace Abstract: Graph Foundation Models (GFMs) on text-attributed graphs (TAGs) align graph representations with language semantics to support transferable graph learning. Despite these advantages, the backdoor vulnerability of GFMs on TAGs remains insufficiently ...

📖 Read original article


250. Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data ​

Author: Renfei Zhang, Niloofar Mireshghallah
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21727v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) is deployed to make models better at reasoning tasks, but its side effect on what models will divulge is under studied. Here we show that RLVR on facts increases extraction of personally identif...

📖 Read original article


251. DAW: Dynamics-Aware Weighting for Deep Learning Forecasts of Chaotic Systems ​

Author: Zhou Fang, Gianmarco Mengaldo
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph

arXiv:2608.22277v2 Announce Type: replace Abstract: Deep learning surrogates for forecasting chaotic dynamical systems suffer from catastrophic error accumulation over long-term autoregressive rollouts. This behavior is partly tied to the underlying systems: chaotic spatiotemporal systems, such as t...

📖 Read original article


252. Beyond Dense Adam States: Adaptive Log-Space Quantization for Memory-Efficient Optimizers ​

Author: Yan Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.22322v2 Announce Type: replace Abstract: Low-precision optimizer-state methods are commonly designed and evaluated for dense Adam-style first and second moments. Memory-efficient optimizers depart from this setting: Adafactor factorizes second moments, CAME adds factored confidence states...

📖 Read original article


253. Clinical Graph-JEPA: Predictive Patient-State Knowledge Graphs for Cognitive Decision Support ​

Author: Kushagra Yadav, Nalin Prabhath, Amit Lamba, James E. Schrager, Goeun Han, Yining Mao
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22583v2 Announce Type: replace Abstract: Clinical records contain rich evidence about patient state, but converting that evidence into reliable, structured knowledge graphs remains difficult because extraction errors, ontology mismatch, missing relations, and temporal ambiguity can propag...

📖 Read original article


254. FedCC: Towards Addressing Label Distribution Skews in Distillation-Based Federated Learning ​

Author: Wenxuan Ye, Onur Ayan, Xueli An, Georg Carle
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.23031v2 Announce Type: replace Abstract: Federated Learning (FL) enables distributed clients to collaboratively train models without sharing raw data, making it promising for leveraging massive devices in communication networks. In distillation-based FL, each client applies its local mode...

📖 Read original article


255. Beyond Point Predictions: Uncertainty-Aware Satellite Poverty Mapping for Public Policy ​

Author: Markus B. Pettersson, James Bailie, Mohammad Kakooei, Eagon Meng, Adel Daoud
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23322v2 Announce Type: replace Abstract: Despite their critical importance for policy and research, high-resolution poverty data remain limited across much of Africa. Machine learning (ML) with earth observation (EO) imagery has recently emerged as a way to supplement these data by predic...

📖 Read original article


256. A Theory of Speciation in Generative Diffusion Models on Compact Riemannian Manifolds ​

Author: Alessio Marta, Paola Causin
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23798v2 Announce Type: replace Abstract: Speciation in generative diffusion models denotes the emergence of distinct stable branches during denoising, through which initially undifferentiated trajectories progressively commit to different data classes. In this work we develop an intrinsic...

📖 Read original article


257. JEPA-x: Cross-Predictive Physics Grounding for Forecastable Latent Dynamics ​

Author: Kehan Wen, Ziming Li, Siyuan Luo, Fan Shi
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.24044v2 Announce Type: replace Abstract: Latent world models plan by predicting how candidate actions advance learned latent dynamics. In self-predictive models, however, the encoder and predictor are optimized jointly and can co-adapt to latent transitions that are easy to predict but we...

📖 Read original article


258. IAPO: Influence-Aware Policy Optimization for Credit Assignment in Multi-Turn Service Agents ​

Author: Bo Ren, Yirong Mao, Yi Yang, Wenhui Que
Published: 8/27/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.24588v2 Announce Type: replace Abstract: Large Language Model (LLM) agents increasingly solve long-horizon tasks through multi-turn interactions with users and external tools. In these settings, relevant task information often unfolds over time rather than being fully specified at the ini...

📖 Read original article


259. Advancements in Content-Based Image Retrieval: A Comprehensive Survey of Relevance Feedback Techniques ​

Author: Hamed Qazanfari, Mohammad M. AlyanNezhadi, Zohreh Nozari Khoshdaregi
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.IR, cs.LG

arXiv:2312.10089v2 Announce Type: replace-cross Abstract: Content-based image retrieval (CBIR) systems have emerged as crucial tools in the field of computer vision, allowing for image search based on visual content rather than relying solely on metadata. This survey paper presents a comprehensive o...

📖 Read original article


260. Generative Modeling by Minimizing the Wasserstein-2 Loss ​

Author: Yu-Jui Huang, Zachariah Malik
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2406.13619v5 Announce Type: replace-cross Abstract: This paper develops a generative model by minimizing the second-order Wasserstein loss (the $W_2$ loss) through a distribution-dependent ordinary differential equation (ODE), whose dynamics involves the Kantorovich potential associated with t...

📖 Read original article


261. Infer Human's Intentions Before Following Natural Language Instructions ​

Author: Yanming Wan, Yue Wu, Yiping Wang, Jiayuan Mao, Natasha Jaques
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2409.18073v2 Announce Type: replace-cross Abstract: For AI agents to be helpful to humans, they should be able to follow natural language instructions to complete everyday cooperative tasks in human environments. However, real human instructions inherently possess ambiguity, because the human ...

📖 Read original article


262. Non-Asymptotic Bounds for Closed-Loop Identification of Sub-Exponentially Growing Nonlinear Stochastic Systems ​

Author: Seth Siriya, Jingge Zhu, Dragan Ne\v{s}i'c, Ye Pu
Published: 8/27/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC

arXiv:2412.04157v2 Announce Type: replace-cross Abstract: We investigate the problem of least squares parameter estimation from single-trajectory data for discrete-time, unstable, closed-loop nonlinear stochastic systems. Specifically, we consider nonlinear systems with linearly parametrised uncerta...

📖 Read original article


263. Generative Modeling: A Review ​

Author: Maria Nareklishvili, Nick Polson, Vadim Sokolov
Published: 8/27/2026, 4:00:00 AM
Categories: stat.CO, cs.LG

arXiv:2501.05458v3 Announce Type: replace-cross Abstract: We organize the generative-modeling literature around three classes of generators, corresponding to three distinct inferential tasks: estimating counterfactual outcome distributions in causal inference, recovering posteriors from simulated pa...

📖 Read original article


264. Thermodynamic cost of inference and learning in physical neural networks ​

Author: Alexei V. Tkachenko
Published: 8/27/2026, 4:00:00 AM
Categories: cond-mat.stat-mech, cs.LG, physics.data-an, q-bio.NC

arXiv:2503.09980v4 Announce Type: replace-cross Abstract: How much of the energy consumed by artificial neural networks is set by physics rather than by implementation? For irreversible digital hardware the reference is Landauer's principle, which charges $k_B T\ln 2$ per erased bit. We map a generi...

📖 Read original article


265. Gradient-based Sample Selection for Faster Bayesian Optimization ​

Author: Qiyu Wei, Haowei Wang, Zirui Cao, Songhao Wang, Richard Allmendinger, Mauricio A 'Alvarez
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2504.07742v4 Announce Type: replace-cross Abstract: Bayesian optimization (BO) is an effective technique for black-box optimization. However, its applicability is typically limited to moderate-budget problems due to the cubic complexity of fitting the Gaussian process (GP) surrogate model. In ...

📖 Read original article


266. Evolutionary chemical learning in dimerization networks ​

Author: Alexei V. Tkachenko, Bortolo Matteo Mognetti, Sergei Maslov
Published: 8/27/2026, 4:00:00 AM
Categories: cond-mat.stat-mech, cond-mat.dis-nn, cs.LG, nlin.AO, physics.data-an, q-bio.MN

arXiv:2506.14006v2 Announce Type: replace-cross Abstract: We present a framework for chemical learning based on Competitive Dimerization Networks (CDNs) - systems in which multiple molecular species, e.g., proteins, DNA oligomers, or RNA oligomers, reversibly bind to form dimers. We show numerically...

📖 Read original article


267. AI/ML Life Cycle Management for Interoperable AI Native RAN ​

Author: Chu-Hsiang Huang, Yuan-Chih Fan Chiang, Chao-Kai Wen, Geoffrey Ye Li
Published: 8/27/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT

arXiv:2507.18538v3 Announce Type: replace-cross Abstract: Artificial intelligence (AI) and machine learning (ML) are rapidly becoming integral to the 5G Radio Access Network (RAN), enabling beam management, channel state information (CSI) feedback, positioning, and mobility prediction. However, with...

📖 Read original article


268. Model-Agnostic Open-Set Air-to-Air Visual Object Detection for Reliable UAV Perception ​

Author: Spyridon Loukovitis, Anastasios Arsenos, Vasileios Karampinis, Athanasios Voulodimos
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO

arXiv:2509.09297v2 Announce Type: replace-cross Abstract: Open-set detection is crucial for robust UAV autonomy in air-to-air object detection under real-world conditions. Traditional closed-set detectors degrade significantly under domain shifts and flight data corruption, posing risks to safety-cr...

📖 Read original article


269. Three-Way Open-Set Detection for Robust Autonomous Navigation ​

Author: Spyridon Loukovitis, Vasileios Karampinis, Athanasios Voulodimos
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO

arXiv:2511.15343v2 Announce Type: replace-cross Abstract: Autonomous navigation in complex scenes requires reliable perception across scenarios that the model did not encounter during its training. Along its route, an autonomous framework encounters objects it was trained to recognize, obstacles it ...

📖 Read original article


270. Ladder Up, Memory Down: Low-Cost Fine-Tuning With Side Nets ​

Author: Estelle Zheng, Nathan Cerisara, S'ebastien Warichet, Emmanuel Helbert, Christophe Cerisara
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2512.14237v2 Announce Type: replace-cross Abstract: Fine-tuning large language models (LLMs) is often limited by the memory available on commodity GPUs. Parameter-efficient fine-tuning (PEFT) methods such as QLoRA reduce the number of trainable parameters, yet still incur high memory usage ind...

📖 Read original article


271. Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations ​

Author: Qianli Wang, Nils Feldhus, Pepa Atanasova, Fedor Splitt, Simon Ostermann, Sebastian M"oller, Vera Schmitt
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2601.00282v3 Announce Type: replace-cross Abstract: Quantization is widely used to accelerate inference and streamline the deployment of large language models (LLMs), yet its effects on self-explanations (SEs) remain unexplored. SEs, generated by LLMs to justify their own outputs, require reas...

📖 Read original article


272. iFlip: Iterative Feedback-driven Counterfactual Example Refinement ​

Author: Yilong Wang, Qianli Wang, Nils Feldhus
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2601.01446v2 Announce Type: replace-cross Abstract: Counterfactual examples are minimal edits to an input that alter a model's prediction. They are widely employed in explainable AI to probe model behavior and in natural language processing (NLP) to augment training data. However, generating v...

📖 Read original article


273. Generalized Riesz Regression: A Unified Framework for Debiased Machine Learning with Riesz Representer Fitting under Bregman Divergence ​

Author: Masahiro Kato
Published: 8/27/2026, 4:00:00 AM
Categories: econ.EM, cs.LG, math.ST, stat.ME, stat.ML, stat.TH

arXiv:2601.07752v4 Announce Type: replace-cross Abstract: Estimating the Riesz representer is central to debiased machine learning, yet the generator and representer model determine which regression directions their first-order conditions protect. We introduce generalized Riesz regression, which min...

📖 Read original article


274. Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing ​

Author: Dohun Lee, Chun-Hao Paul Huang, Xuelin Chen, Jong Chul Ye, Duygu Ceylan, Hyeonho Jeong
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2601.16296v3 Announce Type: replace-cross Abstract: Video-to-video diffusion models achieve impressive single-turn editing performance, but practical editing workflows are inherently iterative. When edits are applied sequentially, existing models treat each turn independently, often causing pr...

📖 Read original article


275. Performance uncertainty in medical image analysis: a large-scale investigation of confidence intervals ​

Author: Pascaline Andr'e (Sorbonne Universit'e, Institut du Cerveau - Paris Brain Institute - ICM, CNRS, Inria, Inserm, AP-HP, H^opital de la Piti'e-Salp^etri`ere, Paris, France), Charles Heitz (Sorbonne Universit'e, Institut du Cerveau - Paris Brain Institute - ICM, CNRS, Inria, Inserm, AP-HP, H^opital de la Piti'e-Salp^etri`ere, Paris, France), Evangelia Christodoulou (German Cancer Research Center), Annika Reinke (German Cancer Research Center), Carole H. Sudre (Unit for Lifelong Health and Ageing at UCL, Department of Population Science and Experimental Medicine and Hawkes InstituteCentre for Medical Image Computing, Department of Computer Science, University College London, UK), Michela Antonelli (School of Biomedical Engineering and Imaging Science, King's College London, UK), Patrick Godau (German Cancer Research Center), M. Jorge Cardoso (School of Biomedical Engineering and Imaging Science, King's College London, UK), Antoine Gilson (Sorbonne Universit'e, Institut du Cerveau - Paris Brain Institute - ICM, CNRS, Inria, Inserm, AP-HP, H^opital de la Piti'e-Salp^etri`ere, Paris, France), Sophie Tezenas du Montcel (Sorbonne Universit'e, Institut du Cerveau - Paris Brain Institute - ICM, CNRS, Inria, Inserm, AP-HP, H^opital de la Piti'e-Salp^etri`ere, Paris, France), Ga"el Varoquaux (SODA project team, Inria Saclay-^Ile-de-France, France), Lena Maier-Hein (German Cancer Research Center), Olivier Colliot (Sorbonne Universit'e, Institut du Cerveau - Paris Brain Institute - ICM, CNRS, Inria, Inserm, AP-HP, H^opital de la Piti'e-Salp^etri`ere, Paris, France)
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2601.17103v2 Announce Type: replace-cross Abstract: Performance uncertainty quantification is essential for reliable validation and eventual clinical translation of medical imaging artificial intelligence (AI). Confidence intervals (CIs) play a central role in this process by indicating how pr...

📖 Read original article


276. Edge-Local and Qubit-Efficient Quantum Graph Learning for the NISQ Era ​

Author: Armin Ahmadkhaniha, Jake Doliskani
Published: 8/27/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG

arXiv:2602.16018v3 Announce Type: replace-cross Abstract: Graph neural networks (GNNs) are a powerful framework for learning representations from graph-structured data, but their direct implementation on near-term quantum hardware remains challenging due to circuit depth, multi-qubit interactions, a...

📖 Read original article


277. Quantum Scrambling Born Machine ​

Author: Marcin P{\l}odzie'n
Published: 8/27/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2602.17281v2 Announce Type: replace-cross Abstract: Quantum generative modeling, where the Born rule naturally defines probability distributions through measurement of parameterized quantum states, is a promising near-term application of quantum computing. We propose a Quantum Scrambling Born ...

📖 Read original article


278. ST-Lite: Training-Free KV Cache Compression with Spatio-Trajectory Guidance for Long-Horizon GUI Agents ​

Author: Bowen Zhou, Zhou Xu, Wanli Li, Jingyu Xiao, Pingan Gan, Haoqian Wang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2603.00188v3 Announce Type: replace-cross Abstract: Training-free KV cache compression is essential for deploying vision-language GUI agents under memory and latency constraints, yet existing methods are designed for generic language workloads and ignore the distinctive structure of GUI intera...

📖 Read original article


279. TTSR: Test-Time Self-Evolving via Reflection ​

Author: Haoyang He, Zihua Rong, Yunjia Zhao, Lan Yang, Jian Chang, Honggang Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2603.03297v2 Announce Type: replace-cross Abstract: Test-time training (TTT) adapts large language models (LLMs) during inference using only unlabeled test inputs. Existing methods, however, face two major bottlenecks on hard reasoning tasks: (1) \emph{lack of learnable samples}, as self-gener...

📖 Read original article


280. OpenSanctions Pairs: Large-Scale Entity Matching with LLMs ​

Author: Chandler Smith, Magnus Sesodia, Friedrich Lindenberg, Christian Schroeder de Witt
Published: 8/27/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG

arXiv:2603.11051v2 Announce Type: replace-cross Abstract: We release OpenSanctions Pairs, the first large-scale public benchmark for entity matching on sanctions and OSINT data. The dataset includes 755,540 expert-labeled pairs over 1 million entities, aggregated from 293 source datasets across 45 j...

📖 Read original article


281. Regularized Latent Dynamics Prediction is a Strong Baseline For Behavioral Foundation Models ​

Author: Pranaya Jajoo, Harshit Sikchi, Siddhant Agarwal, Amy Zhang, Scott Niekum, Martha White
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.RO

arXiv:2603.15857v2 Announce Type: replace-cross Abstract: Behavioral Foundation Models (BFMs) produce agents with the capability to adapt to any unknown reward or task. These methods, however, are only able to produce near-optimal policies for the reward functions that are in the span of some pre-ex...

📖 Read original article


282. Inverse Design of Inorganic Compounds with Generative AI ​

Author: Hannes Kneiding, Luc'ia Mor'an-Gonz'alez, Nishamol Kuriakose, Ainara Nova, David Balcells
Published: 8/27/2026, 4:00:00 AM
Categories: physics.chem-ph, cond-mat.mtrl-sci, cs.LG

arXiv:2604.11827v3 Announce Type: replace-cross Abstract: Machine learning is revolutionizing chemistry. Beyond the value of predictive models accelerating virtual screening, generative AI aims at enabling inverse design, reversing the compound-to-property prediction paradigm into property-to-compou...

📖 Read original article


283. Cross-Domain Transfer with Particle Physics Foundation Models: From Jets to Neutrino Interactions ​

Author: Gregor Krzmanc, Vinicius Mikuni, Benjamin Nachman, Callum Wilkinson
Published: 8/27/2026, 4:00:00 AM
Categories: hep-ex, cs.LG, hep-ph, physics.data-an

arXiv:2604.12364v2 Announce Type: replace-cross Abstract: Future AI-based studies in particle physics will likely start from a foundation model to accelerate training and enhance sensitivity. As a step toward a general-purpose foundation model for particle physics, we investigate whether the OmniLea...

📖 Read original article


284. Decomposing Gradient Suppression in Barren Plateaus: Activity, Sign Organization, and Coupling ​

Author: Pilsung Kang
Published: 8/27/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2605.01319v2 Announce Type: replace-cross Abstract: Barren plateaus (BPs) are conventionally characterized by suppressed gradient variance, but this aggregate description does not reveal how the loss of gradient signal is composed across Hamiltonian terms. We introduce a term-resolved framewor...

📖 Read original article


285. Reconstruction of Personally Identifiable Information from Proprietary Data in Supervised Fine-Tuned Models ​

Author: Sae Furukawa, Alina Oprea
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG

arXiv:2605.12264v2 Announce Type: replace-cross Abstract: Supervised Finetuning (SFT) has become one of the primary methods for adapting a large language model (LLM) with extensive pre-trained knowledge to domain-specific, instruction-following tasks. SFT datasets, composed of instruction-response p...

📖 Read original article


286. Minimalist Visual Inertial Odometry ​

Author: Francesco Pasti, Jeremy Klotz, Nicola Bellotto, Shree K. Nayar
Published: 8/27/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2605.19990v2 Announce Type: replace-cross Abstract: Visual-Inertial Odometry (VIO), which is critical to mobile robot navigation, uses cameras with a large number of pixels. Capturing and processing camera images requires significant resources. This work presents a minimalist approach to plana...

📖 Read original article


287. Optimal Design for Multinomial Logit Model with Applications to Best Assortment Identification ​

Author: Joongkyu Lee, Min-hwan Oh
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2605.25592v2 Announce Type: replace-cross Abstract: We study optimal experimental design for multinomial logit (MNL) bandits, where an agent repeatedly selects a subset of $K$ items from a ground set of size $N$ and observes single-choice feedback. Unlike linear or generalized linear bandits, ...

📖 Read original article


288. Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification ​

Author: Dylan Bouchard, Mohit Singh Chauhan, Zeya Ahmad, Ho-Kyeong Ra
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2605.28500v2 Announce Type: replace-cross Abstract: Large language models have shown impressive capabilities in code generation, yet they often produce functionally incorrect code. Uncertainty quantification (UQ) methods have emerged as a promising approach for detecting hallucinations in natu...

📖 Read original article


289. Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense ​

Author: Shuhao Zhang, Jiarui Li, Qi Cao, Ruiyi Zhang, Pengtao Xie
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2605.30837v3 Announce Type: replace-cross Abstract: Prompt-injection detectors are heterogeneous: each is strong on a different slice of attacks, and none is always reliable. Yet existing systems still treat detection as a fixed single-detector pipeline, committing every request to one detecto...

📖 Read original article


290. Online Pandora's Box for Contextual LLM Cascading ​

Author: Alexandre Belloni, Yan Chen, Yehua Wei
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, econ.EM, stat.ML

arXiv:2606.07392v2 Announce Type: replace-cross Abstract: Motivated by Large Language Model (LLM) cascading, we propose an online contextual Pandora's Box model for adaptively querying and selecting LLM APIs. In each period, a decision-maker observes a request context and faces a two-phase decision ...

📖 Read original article


291. When Probing Accuracy Saturates, Fragility Resolves: A Complementary Metric for LLM Pre-Training Analysis ​

Author: Orion Reblitz-Richardson
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2606.11375v2 Announce Type: replace-cross Abstract: Standard linear probing declares a property "encoded" when a classifier on hidden states achieves high accuracy. The protocol works well on a snapshot but breaks across pre-training, with probe accuracy saturating within the first few thousan...

📖 Read original article


292. ClayBuddy: A Framework, Evaluation, & Mitigation of Coding Agent Failures ​

Author: Kenneth Ge, Andre Assis
Published: 8/27/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2606.19380v5 Announce Type: replace-cross Abstract: Widespread deployment of AI agents in software engineering is surfacing a long tail of rare but highly dangerous misalignment bugs. Since sampling this behavior is intractable, we decompose these failures into three distinct mechanisms: under...

📖 Read original article


293. Learning to Prompt: Improving Student Engagement with Adaptive LLM-based High-School Tutoring ​

Author: Po-Chin Chang, Nicholas Hogan, Aske Plaat, Michiel T. van der Meer
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.HC, cs.LG

arXiv:2606.20138v2 Announce Type: replace-cross Abstract: LLMs can personalize education, although current static-prompt tutoring systems struggle to adapt to diverse academic disciplines. We develop and test a system with subject-aware prompting, based on 14 pedagogical features (e.g., tutor scaffo...

📖 Read original article


294. RoboMME-Interference: Benchmarking Robot Memory Under Interference ​

Author: Soumil Rathi
Published: 8/27/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2606.22338v3 Announce Type: replace-cross Abstract: Robots deployed in realistic settings will accumulate experience across many sessions and tasks over their deployment. The robot's tasks may often require it to remember information from multiple sessions ago, making long-context robot memory...

📖 Read original article


295. From Idea to Prototype in an Afternoon: Scaffolded, AI-Assisted Rapid VA Prototyping ​

Author: Gennady Andrienko, Natalia Andrienko
Published: 8/27/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.LG

arXiv:2606.31311v2 Announce Type: replace-cross Abstract: Testing a new visual-analytics idea usually takes months: one needs to find a realistic data set, clean it, and implement an interactive prototype. We describe a case where a workflow language and an AI assistant reduced this effort to one af...

📖 Read original article


296. Smooth $\%$MinMax: A Differentiable Relaxation for Codon Harmonization ​

Author: Yoonho Jeong, Hyunwoo Choi, Ryan Fernandez Medina Hariri, Eok Kyun Lee, Seung Seo Lee, Insung S. Choi
Published: 8/27/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG

arXiv:2607.03881v2 Announce Type: replace-cross Abstract: Codon harmonization aims to adapt the coding sequences for heterologous expression while preserving the native-like patterns of frequent and rare codons that may influence local translation dynamics and co-translational protein folding. Howev...

📖 Read original article


297. Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents ​

Author: Amirmohammad Farzaneh, Osvaldo Simeone
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.IT, cs.LG, math.IT

arXiv:2607.26865v2 Announce Type: replace-cross Abstract: LLM agents following the ReAct paradigm are promising enablers of complex multi-step tasks, including multi-hop question answering, code generation, and control of physical AI systems. Yet, when deployed at the edge, they must tightly manage ...

📖 Read original article


298. Analytical and Bootstrap Confidence Intervals of Double Machine Learning: Simulation studies and an application to rural-urban difference in obesity prevalence ​

Author: Haozheng Xu, Siyuan Ma, Qingyan Xiang
Published: 8/27/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP

arXiv:2607.29456v2 Announce Type: replace-cross Abstract: Double Machine Learning (DML) is a popular approach for treatment effect estimation in various settings, which allows a wide range of flexible machine learning methods to be used for nuisance parameter estimation while preserving valid infere...

📖 Read original article


299. LILAC: An Idempotent Neural Speech Codec ​

Author: June Young Yi, Dongwook Lee, Jiheum Yeom, Sungroh Yoon
Published: 8/27/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS

arXiv:2608.05727v2 Announce Type: replace-cross Abstract: Neural Audio Codecs are widely adopted in speech generation and editing. However, existing neural audio codecs are not idempotent: across the paper's twelve baseline systems, every configuration tested rewrites, on average, at least 15% of it...

📖 Read original article


300. Privileged Likelihood Is Not Automatically Value: Three Checks for Token Credit in On-Policy Self-Distillation ​

Author: Xuan-Phi Nguyen, Zeyu Leo Liu, Yang Li, Shrey Pandit, Yiran Zhao, Anurag Koul, Shafiq Joty
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.09263v2 Announce Type: replace-cross Abstract: On-policy self-distillation aims to improve upon reinforcement learning from verifiable rewards (RLVR) by providing token-level scores derived from privileged information, such as reference solutions or critic feedback. These scores are treat...

📖 Read original article


301. Gated Recurrent Transformers: Expressive Depth through Recurrent Modulation ​

Author: Amr Hegazy, Amr Alanwar, Mostafa Elhoushi
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.15062v4 Announce Type: replace-cross Abstract: Scaling transformer language models creates an inherent tension between expressivity and memory efficiency. While unique weights across layers preserve functional specialization---from input-grounding to abstract refinement---they incur a sub...

📖 Read original article


302. Offline Reinforcement Learning for Hemodynamic Management of Sepsis in the ICU: a MIMIC-IV Study with Dual Off-Policy Evaluation ​

Author: Marc P'erez-Roig, David Fern'andez-Narro, Carlos S'aez
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.16482v2 Announce Type: replace-cross Abstract: The dosing of intravenous fluids and vasopressors in sepsis is a sequential decision made under uncertainty and guided largely by clinical judgment, which makes it a natural target for reinforcement learning from historical care. Because a le...

📖 Read original article


303. TokEval: A Tokenizer Evaluation Suite ​

Author: Clara Meister
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.18062v2 Announce Type: replace-cross Abstract: Language model tokenizers are typically selected with minimal evaluation, despite the fact that their design choices directly impact model capabilities. This can be partly attributed to a limited understanding of which tokenizer properties af...

📖 Read original article


304. EDGE: Experience-Distillation for Guided Exploration in Agentic Reinforcement Learning ​

Author: Can Xie, Yuyi Zhou, Wen Yang, Ziyi zhang, Siyao Song, Yingzhuo Deng, Shuo Ren, Jiajun Zhang
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.21946v2 Announce Type: replace-cross Abstract: Reinforcement learning with outcome-based objectives such as GRPO enables LLM-based agents to solve complex, long-horizon tasks, yet the reusable exploration patterns embedded in interaction trajectories are largely discarded after a single p...

📖 Read original article


305. DELE-w0.5: Inferring Action from Future Latent State for Robotic Manipulation ​

Author: Fenghao Lei, Zhixiong Huang, Long Yang, Jiabao Chen, Peilin Huang, Han Fu, Zhuo Li, Xiaoxue Ren
Published: 8/27/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG

arXiv:2608.22067v3 Announce Type: replace-cross Abstract: World-Action Models (WAMs) build robot control on video-generation backbones, which jointly predict dense future visual trajectories and robot actions. We argue that video generation is an unnecessary intermediate objective for world-action m...

📖 Read original article


306. Autonomous Cyber Defense: Real-Time Attack Detection and Mitigation in Software-Defined Networks Using Machine Learning ​

Author: Alexandre Amaral, Fernando Moro, Ana Malheiro
Published: 8/27/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.22075v3 Announce Type: replace-cross Abstract: Autonomous response has evolved into a timing-critical challenge rather than solely a matter of detection accuracy. In recent intrusions, the interval between initial access and the first lateral movement has been observed to be as short as 2...

📖 Read original article


307. TailSieve: Partial-Rollout-Guided Tail Routing for LLM Rollouts ​

Author: Tianqi Xu, Lu Lv, Haoyang Huang, Wenjie Huang, Zhanming Shen, Yuhao Shen, Baolin Zhang, Xinyi Hu, Shuang Ge, Jun Dai, Tianyu Liu, Suorong Yang, Zhikai Li, Ye Bai, Jun Zhang, Lei Chen, Yue Li, Mingchen Wan
Published: 8/27/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.22788v2 Announce Type: replace-cross Abstract: Large-scale rollouts have become a core component of modern LLM systems, spanning reinforcement learning (RL) post-training, on-policy distillation (OPD), and sampling-heavy evaluation pipelines. Unlike online serving, which is typically opti...

📖 Read original article


308. Learning to Act While Waiting: RL Finetuning of Generalist Robot Policies Under Inference Latency ​

Author: Brian Zhu, Momen Khalil, E Harrison, Emanuele Poggi, Philipp Schmitt, Bernd Kast, Philine Meister, Pranav Atreya, Qiyang Li, Finn Ferchau, Cesar Colmenero, Yash Shahapurkar, Gokul Narayanan, Melih Erdogan, Kai Wurm, Georg von Wichert, Oier Mees, Eugen Solowjow, Andrew Wagenmaker, Sergey Levine
Published: 8/27/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.23831v2 Announce Type: replace-cross Abstract: While reinforcement learning (RL) allows generalist robot policies to continually improve during deployment, the large model size of modern generalist policies, such as VLAs, poses a fundamental obstacle to effective RL improvement. In partic...

📖 Read original article