Skip to content

arXiv cs.LG - 2026-08-20 ​

237 items collected.


1. Proactive Road Safety Intervention in Australia: Predicting Risky Driving Hotspots from Connected Vehicle Data ​

Author: Adriana-Simona Mih\u{a}i\c{t}\u{a}, Clarence Cheung, Artur Grigorev, Tuo Mao, David Lillo-Trynes
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2608.16913v1 Announce Type: new Abstract: Road safety monitoring has historically been reactive, relying on crash-record analysis after fatalities and injuries have already occurred. Proactive identification of high-risk locations and dangerous driving behaviour before incidents occur is a cri...

📖 Read original article


2. Detecting and Discriminating Operator Misspecification in Hybrid PDE-Parameter Learning: a Reference-Free Instrument, with Discrimination Bounded In Sample ​

Author: Eric Fock
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, stat.ME

arXiv:2608.16925v1 Announce Type: new Abstract: We build an instrument that reads, from a single fit and with no oracle, whether the operator a hybrid PDE-parameter estimator postulates is wrong-and separates that from a merely unidentifiable parameter. On one self-adjoint parabolic inverse problem,...

📖 Read original article


3. Data-DPO: Direct Preference Optimization for Target Model Data Selection in LLM Post-Training ​

Author: Peng Sun, Yi Yang, Antong Zhang, Chunxiao Li, Yanbo Wang, Dianbo Liu, xin chen, Kai Yu, Lu Chen, Tianfan Fu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16926v1 Announce Type: new Abstract: Data selection in supervised fine-tuning aims to select a small set of effective samples from large-scale candidate data, reducing training cost while preserving model performance. However, existing methods usually treat data value as a relatively stat...

📖 Read original article


4. Hierarchical Data Selection via Manifold Coverage and Sparse Feature Coverage in LLM Post-training ​

Author: Peng Sun, Yi Yang, Antong Zhang, Chunxiao Li, Yanbo Wang, Dianbo Liu, xin chen, Kai Yu, Lu Chen, Tianfan Fu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.16927v1 Announce Type: new Abstract: As supervised fine-tuning data continues to scale, selecting high-value subsets from large candidate pools is crucial for reducing training cost and improving model performance. Existing methods often measure diversity directly in the original embeddin...

📖 Read original article


5. Benchmarking Classical and Transformer-Based Models for Document Sensitivity Classification ​

Author: Aleesha Zainab, Muhammad Ahmed Khalid, Faheem Ullah Khan, Asifullah Khan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16928v1 Announce Type: new Abstract: Automatic sensitivity classification of organizational documents is a critical yet underserved problem, where the consequences of misclassification range from regulatory violations to security breaches. While AI-based approaches offer a scalable altern...

📖 Read original article


6. Mr.Dec: Daily-Scale Longitudinal Multimodal Modeling for 30-Day Readmission Prediction ​

Author: Minjun Kim, Jong Hak Moon
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16929v1 Announce Type: new Abstract: Predicting 30-day hospital readmission is essential for assessing patient stability and optimizing healthcare resources. As clinical risk evolves with the accumulation of evidence during hospitalization, capturing these dynamic trajectories is essentia...

📖 Read original article


7. EMAN: Optimization-Driven Capacity Growth through Path Emergence in Multi-Task Learning ​

Author: Chenlei Fang, Jingchen Li, Hongzong LI, Qingyao Li, Yixuan Zhang, Huarui Wu, Haobin Shi, Chunjiang Zhao
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16930v1 Announce Type: new Abstract: Existing multi-task learning methods rely on hard sharing, multiple paths or experts, adaptive sharing, and dynamic expansion. However, their capacity changes are usually constrained by predefined structures or triggered by task boundaries and conflict...

📖 Read original article


8. SW-ProxyCE: Zero-Query Adversarial Transfer from Public EEG Encoders to Private Downstream Models ​

Author: Linhua Cong, Dingkun Liu, Dongrui Wu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16931v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models have recently emerged as a promising paradigm for EEG decoding by learning reusable representations from large-scale heterogeneous neural recordings. However, the open release of EEG foundation encoders, w...

📖 Read original article


9. DOW-KE: Anchor-Free Multi-Layer Knowledge Editing via Direct End-to-End Weight Optimization ​

Author: Ran Chen, Junbo Zhang, Qianli Zhou, Xinyang Deng, Wen Jiang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16932v1 Announce Type: new Abstract: Multi-layer locate-then-edit methods for knowledge editing first optimize target residual-stream activations (anchors) at selected layers, then realize them layer by layer as weight updates. This pipeline optimizes an intermediate representation but de...

📖 Read original article


10. Study-Strategy Clusters from EdNet Logs Track Engagement, Not Mastery ​

Author: Qingchuan Lyu, Yingxin Li, Albert Yang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CY, stat.AP

arXiv:2608.16963v1 Announce Type: new Abstract: Learning analytics often treats unsupervised clusters of intelligent tutoring system (ITS) logs as learner types that should predict learning. We test that assumption on EdNet-KT3. Clustering study-strategy features (resource use, revision, video, prob...

📖 Read original article


Author: A. Rahaman, A. Quadir, M. Tanveer
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16965v1 Announce Type: new Abstract: The dominance of majority classes in real-world datasets poses a fundamental challenge to randomized neural networks, often biasing decision boundaries and overlooking critical minority samples. Existing remedies, such as synthetic minority over-sampli...

📖 Read original article


12. MultiSigBERT: Beyond Survival Analysis through Multimodal and Sequential Modeling in Oncology ​

Author: Paul Minchella, St'ephane Chr'etien, Guillaume Metzler, Lo"ic Verlingue, R'emi Vaucher
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16972v1 Announce Type: new Abstract: Machine learning has become an essential component of modern healthcare, where the integration of heterogeneous data sources offers unprecedented opportunities to improve clinical decision-making. Electronic Health Records (EHR) contain complementary i...

📖 Read original article


13. Position: Fairness Failure in Generative Models is an Evaluation Problem ​

Author: Mariia Vladimirova, Jean-Yves Franceschi, Thibaut Issenhuth
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16974v1 Announce Type: new Abstract: Despite groundbreaking advancements in generative models during the last decade, concerns about their lack of fairness, reinforcing societal inequalities and harming marginalized groups, remain under-addressed and difficult to act upon. This position p...

📖 Read original article


14. Agents unlock new capabilities through Switching LoRA Adapters as a Tool (SLAaaT) ​

Author: Kenneth Ge
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17034v1 Announce Type: new Abstract: Post-training can unlock new capabilities and improve performance on specialized tasks, but sometimes at the cost of catastrophic forgetting in other domains. This poses a problem in long agent trajectories that compose different capabilities. We rejec...

📖 Read original article


15. J-Miner: Recovering Executable Decision Knowledge from Language-Model Classifiers ​

Author: Yunfan Gao, Xinyi Huang, Tao Sheng, Haorui Song, Yun Xiong, Haofen Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.17063v1 Announce Type: new Abstract: Large language models can be fine-tuned into specialized classifiers that perform well across diverse text tasks and make complex judgments, but they typically expose only final labels, leaving the decision knowledge acquired through fine-tuning implic...

📖 Read original article


16. Certified but Private: Scalable Zero-Knowledge Proofs for Neural Network Guarantees ​

Author: Youwei Zhong, Ben Merbaum, Timos Antonopoulos, Ning Luo, Charalampos Papamanthou, Katerina Sotiraki, Ruzica Piskac
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.LO

arXiv:2608.17070v1 Announce Type: new Abstract: With the growing deployment of machine learning models, formal guarantees of the robustness and fairness of these models have become increasingly important in safety-critical and legal-compliance settings. However, model parameters are often commercial...

📖 Read original article


17. Dynamic Regime-Aware Conformal Calibration for Reliable Economic Forecast Intervals under Multiple Distribution Shifts ​

Author: Bogdan Oancea
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17079v1 Announce Type: new Abstract: Conformal prediction provides distribution-free prediction intervals but relies on exchangeability, an assumption often violated in economic forecasting because of covariate shift, concept drift, local heterogeneity and latent regimes. We propose Dynam...

📖 Read original article


18. Backward through Time, Algebraically ​

Author: Konstantinos Kogkalidis
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.LO, cs.PL, cs.SY, eess.SY

arXiv:2608.17087v1 Announce Type: new Abstract: Linear temporal logic is a modal extension of propositional logic that allows one to state how a system should behave over time. Its canonical domain is the booleans, but discretely-valued judgements are of little use in steering softly-valued systems ...

📖 Read original article


19. Deep Learning for Cross-Border Electricity Price Forecasting: A Comparative Study ​

Author: Hadeer Elashhab, Sai Srijan Papineni, Marvin Dorn, Veit Hagenmeyer, Benjamin Sch"afer
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17091v1 Announce Type: new Abstract: While publicly available electricity market data presents a valuable resource for forecasting research, the field lacks established benchmark datasets for standardized comparison. As a result, many studies have relied on different datasets and metrics ...

📖 Read original article


20. From Abductive Explanations to Global Logical Rules for Node Classification in SGCs ​

Author: Bryan Lima Cavalcante, Thiago Alves Rocha
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.LO

arXiv:2608.17103v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have achieved remarkable performance in node classification tasks, motivating growing interest in methods capable of explaining their predictions. Recent logic-based approaches, such as LogicXGNN, derive global logical rule...

📖 Read original article


21. Causal Discovery in Equal Variance Linear Gaussian DAGs via SURE-Tuned Ridge Regression ​

Author: Sambit Mishra, Urbashi Mitra
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, stat.ML

arXiv:2608.17132v1 Announce Type: new Abstract: Recovering the directed acyclic graph (DAG) of a structural equation model (SEM) from observational data is a central problem in causal discovery. The iterative gradient descent and per-problem hyperparameter tuning of continuous-optimization methods a...

📖 Read original article


22. Iterative tensor network transformations for element-wise evaluation of elementary and filtering functions ​

Author: Xiao Wang, Tomohiro Hashizume, Pia Siegl, Dieter Jaksch
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, cs.AI, physics.comp-ph, quant-ph

arXiv:2608.17135v1 Announce Type: new Abstract: Tensor networks are powerful formats for compressing large-scale data. However, their application to general data processing has been limited by the difficulty of performing nonlinear operations. Here, we introduce iterative tensor network transformati...

📖 Read original article


23. OraclePhys: A Systematic Framework for LLM Fine-Tuning on Structural Mechanics ​

Author: Mingyu Li, Guorui Song, Jing Lin, Haoqian Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17162v1 Announce Type: new Abstract: What a language model internalizes from fine-tuning is usually diagnosed after the fact. We make it an experimental variable. OraclePhys is a systematic fine-tuning framework with three components: OraclePhys-Bench, an exactly-graded structural-mechani...

📖 Read original article


24. Q-Learning With World Models ​

Author: Perry Dong, Yueru Jia, Chelsea Finn, Dorsa Sadigh
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17163v1 Announce Type: new Abstract: Off-policy reinforcement learning (RL) has become increasingly sample-efficient, enabling applications such as RL fine-tuning of Vision-Language-Action models into reliable, high-performing policies. World models offer a further lever for sample effici...

📖 Read original article


25. SCENARIODIFF: A Scenario-level Guidance Framework for Multimodal Time Series Forecasting--Extended Version ​

Author: Tuan-Binh Tran, Dat Nguyen Cong, Duc-Trong Le, Thanh Trung Huynh, Tung Kieu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17164v1 Announce Type: new Abstract: Textual context such as news, reports, and logs can provide valuable signals for time series forecasting, especially when future dynamics are driven by external events that are not yet visible in historical values. Existing multimodal forecasting metho...

📖 Read original article


26. Population Health-Based Machine Learning Reveals Associations Between Psychosocial Factors and Chronic Kidney Disease ​

Author: Md. Atik Shams, David Eisenberg, Sumaiya Fatema, Asma Sultana, D. M Hasibul Islam, Junnatul Mawa, Anindita Datta, Nafiya Ahmed, Danastan Tasaouf Mridula, SK. Sazid Mahmud, Simon Bin Akter, Tanjila Helaly, Jorge Fresneda Fernandez, Humayera Islam, Tanmoy Sarkar Pias
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17174v1 Announce Type: new Abstract: Chronic kidney disease (CKD) progresses silently and severely undermines quality of life, making early detection critical for improving patient outcomes. We present a two-part study that combines large-scale telehealth data with advanced machine learni...

📖 Read original article


27. Task Specialization Fine-Tuning for Contextual Reinforcement Learning ​

Author: Jianan Zhou, Jung-Hoon Cho, Tianyue Zhou, Han Zheng, Jie Zhang, Roy Dong, Yining Ma, Cathy Wu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17180v1 Announce Type: new Abstract: Contextual Reinforcement Learning (CRL) seeks to generalize classical RL by maximizing task coverage across a context space of related tasks. While prior works often train from scratch and rely on either multi-task learning for a single policy or strat...

📖 Read original article


28. Reinforcement Learning as (Discrete) Potential Theory ​

Author: Christopher Connolly
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.GT

arXiv:2608.17181v1 Announce Type: new Abstract: Reinforcement learning (RL) theory fundamentally depends on probability theory through the Markov chain. There is a deep connection between probability theory and potential theory. This paper reviews that connection and explores the potential-theoretic...

📖 Read original article


29. How smoothing the affinity matrix affects neighborhood preservation in t-SNE ​

Author: Shirin Mohebi, Guillaume Bied, Jefrey Lijffijt
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.17190v1 Announce Type: new Abstract: Dimensionality reduction methods are instrumental to visualize high-dimensional data, and t-SNE stands as one of the most widely used methods due to its emphasis on local neighborhood preservation. A central component of t-SNE is the affinity matrix, w...

📖 Read original article


30. Pessimistic Meta-Induction and Its Limits: Lessons from Frequentist Statistics and Machine Learning Theory ​

Author: Hanti Lin
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ME

arXiv:2608.17213v1 Announce Type: new Abstract: This paper challenges the pessimistic meta-inductive argument against scientific realism by undermining its inductive step rather than its historical premise. Although related challenges already exist, I develop a new one. Drawing on a general epistemo...

📖 Read original article


31. Delta2Gamma: Band-Wise Adaptive Contrastive Learning of EEG for Alzheimer's Disease Detection ​

Author: Chanwoo Park, Chanwoo Kim
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17231v1 Announce Type: new Abstract: Low-cost, scalable screening for dementia remains an open problem. Imaging-based diagnosis is costly and hard to deploy widely. Electroencephalography (EEG) is portable and inexpensive, but its recordings are noisy, vary widely across subjects, and car...

📖 Read original article


32. Physics-Informed and Hybrid Machine Learning in Additive Manufacturing: Application to Fused Filament Fabrication ​

Author: Berkcan Kapusuzoglu, Sankaran Mahadevan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, stat.CO

arXiv:2608.17246v1 Announce Type: new Abstract: This article investigates several physics-informed and hybrid machine learning strategies that incorporate physics knowledge in experimental data-driven deep-learning models for predicting the bond quality and porosity of fused filament fabrication (FF...

📖 Read original article


33. Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL ​

Author: Yunhao Yang, Yuexin Bian, Yunjie Tian, Di Fu, Tianjin Huang, Yuanyuan Shi, Ziang Xiao, Nuno Vasconcelos, Yijiang Li
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.17253v2 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a powerful approach for improving reasoning in language and vision-language models, yet its strongest successes still depend heavily on ground-truth supervision (e.g., verifiable reward). Such annotations are ...

📖 Read original article


34. Understanding Curriculum Learning in Large Language Models via Cross-Difficulty Optimization Dynamics ​

Author: Zhikai Ding, Ziyi Ye
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17268v1 Announce Type: new Abstract: Curriculum learning has been widely adopted in the post-training of large language models by organizing training data from easy to hard. However, its effectiveness varies substantially across reasoning tasks, suggesting that no single curriculum is uni...

📖 Read original article


35. Rethinking Irregular Time Series Forecasting from the Perspective of Basis Functions ​

Author: Rongwen Li, Changjian Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17284v1 Announce Type: new Abstract: Irregular time series forecasting is crucial in many domains, such as healthcare and meteorological observation. However, due to the inherent characteristics of irregular time series, including sparse observations and non-uniform sampling, accurately p...

📖 Read original article


36. Abra: Scaling Diffusion Image Training ​

Author: Kyle Chickering, Wei-An Lin, Swayam Bhanded, Dan Saunders, Akshat Tripathi, Jiaming Song, Shyamal Buch, Xinchen Yan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17286v1 Announce Type: new Abstract: Compute-optimal scaling laws guide the training of frontier language models yet remain largely unexplored for visual generation. We present a systematic scaling law study for text-to-image diffusion models using Abra, a controlled family of flow-matchi...

📖 Read original article


37. Beyond MSE: Rethinking the Evaluation Metric and Benchmarking for Irregular Time Series Forecasting ​

Author: Rongwen Li, Haixin Xie, Xiao Wang, Changjian Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17293v1 Announce Type: new Abstract: Existing research on irregular time-series forecasting has primarily focused on model design, while evaluation metrics remain insufficiently studied. Existing benchmarks typically use mean squared error (MSE) as the evaluation metric. We show that, in ...

📖 Read original article


38. Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements ​

Author: Zhi Zheng, Rongsheng Chen, Yunpeng Ba, Zhenkun Wang, Yee Whye Teh, Wee Sun Lee
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17310v1 Announce Type: new Abstract: Reinforcement Learning (RL) has been promising in single-turn LLM fine-tuning. However, long-horizon agentic reasoning introduces increasingly branching interactions and sparse rewards, exposing several limitations of RL: its heavyweight backpropagatio...

📖 Read original article


39. MoFE: A Novel Mixture-of-Experts Framework with Fourier Neural Operators for Cryptocurrency Forecasting ​

Author: Bowen Liu, Mingming Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17342v1 Announce Type: new Abstract: Forecasting cryptocurrency prices remains a formidable challenge due to inherent non-stationarity, abrupt regime shifts, and multi-scale stochastic dependencies. Conventional deep learning models often struggle to capture complex underlying dynamics, f...

📖 Read original article


40. Tight Bounds for Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function ​

Author: Anh Tuan Nguyen, Viet Anh Nguyen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.17343v1 Announce Type: new Abstract: Data-driven algorithm design frames hyperparameter tuning as a statistical learning problem, but establishing generalization guarantees remains challenging due to the implicit, non-smooth dependence of model performance on hyperparameters. Existing mul...

📖 Read original article


41. Repetition as Reinforcement: Enhancing Sample Efficiency via Instant Episode Repetition in Reinforcement Learning ​

Author: Hoda Yamani, Yuning Xing, Koen van Rijnsoever, Bruce A. MacDonald, Henry Williams
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.17347v1 Announce Type: new Abstract: Repetition is a fundamental mechanism in human learning, where revisiting successful experiences strengthens memory, consolidates skills, and improves future performance. Motivated by this biological principle, we introduce Instant Episode Repetition (...

📖 Read original article


42. CORAM: Coherent Orthogonal Rotation for Model Merging ​

Author: Xinyi Sui, Ziran Liu, Nam Ling, Wei Wang, Wei Jiang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17366v1 Announce Type: new Abstract: Merging finetuned models combines specialized capabilities without joint training or access to the original data. Most methods operate by linear arithmetic in Euclidean weight space, which cannot carry the geometry of the update. Orthogonal Model Mergi...

📖 Read original article


43. Pathology Transport: Optimal-Transport Explanations for Clinical Data, and When Their Heatmaps (Fail to) Localize Disease ​

Author: Lalit Kumar
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17370v1 Announce Type: new Abstract: Generative models promise a route to explainable clinical AI: rather than probe a classifier, model the distributions of healthy and diseased patients and read explanations off the geometry between them. We build such a system - an optimal-transport re...

📖 Read original article


44. Integrating Novelty and Surprise for Experience Prioritization and Exploration in Image-Based Reinforcement Learning ​

Author: Hoda Yamani, Henry Williams, Bruce A. MacDonald
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17373v1 Announce Type: new Abstract: Sample efficiency is a central challenge in reinforcement learning (RL), particularly in image-based domains where agents must learn from high-dimensional visual inputs. Traditional sampling often relies on random or suboptimal experience selection, le...

📖 Read original article


45. GUPO: Gradient Uncertainty-aware Policy Optimization for Post-Training Large Language Models ​

Author: Peizheng Guo, Jianqi Zhang, Xingyu Zhang, Yun Fan, Jiahuan Zhou, Changwen Zheng, Wenwen Qiang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17411v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a widely used approach for post-training Large Language Models (LLMs) for reasoning. In GRPO, the group gradients induced by different queries within the same mini-batch are directly averaged to form...

📖 Read original article


46. General Semantic Knowledge Infusion for Spatio-Temporal Traffic Forecasting ​

Author: Mattis thor Straten, Yannick Wolker, Steffen Strohm, Prathvish Mithare, Ralf Krestel, Matthias Renz
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17440v1 Announce Type: new Abstract: Although Graph Neural Networks (GNNs) have made significant advances in spatio-temporal traffic forecasting, their performance is limited when they rely solely on sensor proximity or road-network topology. This paper presents a spatio-temporal predicti...

📖 Read original article


47. Causal Local States: Scalable Simultaneous Causal Network Inference and Forecasting for Dynamical Systems ​

Author: Jonas Braun, Fabian Fischbach, Daniel K"oglmayr, Sebastian Baur, Christoph R"ath
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17452v1 Announce Type: new Abstract: Machine learning methods predict many real-world systems with remarkable accuracy, but they are typically treated as black boxes that offer no insight into which interactions drive the dynamics. Causal discovery methods reconstruct the interaction netw...

📖 Read original article


48. Evaluating RL Explainability Methods by How Much They Help Fix Bugs in Agents ​

Author: Ram Rachum, Yotam Amitai, B'alint Gyevn'ar, Reuth Mirsky, Cameron Allen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17524v1 Announce Type: new Abstract: This preliminary paper outlines a planned evaluation benchmark for Explainable Reinforcement Learning (XRL) methods. Current evaluations rely on functionally-grounded metrics like faithfulness and compactness, and on human-grounded proxies like subject...

📖 Read original article


49. No Gaussian Required: Contrastive Inverse Dynamics for JEPA World Models ​

Author: Jack Boylan, Chris Hokamp
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17542v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) learn world models by predicting future embeddings, but the objective admits a trivial solution of a constant encoder, so every practical system adds an anti-collapse mechanism (LeCun, 2022; Assran et al...

📖 Read original article


50. Domain-Adapted Molecular Language Models for Efficient Search of Make-on-Demand Libraries ​

Author: Henrik Wille, Luis-Finley Sch"utz, Felix Strieth-Kalthoff
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.17567v1 Announce Type: new Abstract: Pretrained molecular language models are increasingly used as molecular encoders for learning structure-property relationships. However, their practical suitability for molecular discovery within and beyond their pretraining domain remains unclear. Her...

📖 Read original article


51. OOD Detection for EEG-based Machine Learning in High-Risk Environments ​

Author: Philipp Bomatter, Henry Gouk
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17620v1 Announce Type: new Abstract: Machine learning models for electroencephalography (EEG) analysis show great promise across a wide range of applications, but their deployment in high-risk domains is hindered by their vulnerability to distribution shifts. Encountering out-of-distribut...

📖 Read original article


52. rl-triton: High-Performance Triton GPU Kernels for Reinforcement Learning Credit Assignment ​

Author: Lars Simon Zehnder
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.PF

arXiv:2608.17641v2 Announce Type: new Abstract: We present rl-triton, an open-source library of high-performance GPU kernels for reinforcement learning credit assignment, implemented in Triton. The core contribution is a unified associative scan framework that recasts seven distinct RL estimation al...

📖 Read original article


53. Elimination Geometry ​

Author: Mian Huang, Xueqin Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17646v1 Announce Type: new Abstract: This monograph develops elimination geometry (EG), a typed, native-loss, audit-oriented framework for studying when locally optimal objects can be realized by a shared deployment rule. Elimination and compression may erase distinctions required by pred...

📖 Read original article


54. Picard Proximal Monte Carlo for Parallel Bayesian Imaging with Score-Based Generative Priors ​

Author: Deliang Wei, Evan Bell, Wenhan Guo, Yifan Chen, Yu Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17666v1 Announce Type: new Abstract: Bayesian imaging inverse problems often require sampling from high-dimensional posterior distributions. While recent score-based and diffusion models provide expressive Bayesian priors, their sampling procedures remain inherently sequential and computa...

📖 Read original article


55. Conformal Prediction for Molecular Properties under Label Shift ​

Author: Hyeonsu Lee, Juyeon Kim, Erkhembayar Jadamba, Seungjin Choi, Hyunjin Shin
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17678v1 Announce Type: new Abstract: Drug discovery and development underpins healthcare but remains costly and failure-prone. A critical bottleneck lies in predicting molecular properties such as solubility, potency, and toxicity, which directly determine whether a candidate can advance ...

📖 Read original article


56. Cross-View Correspondence Is a Measurement Intervention: Two-Sided Validation for Agent Evaluation and Credit Assignment ​

Author: Zhen Zhang, Ahmad Hafez, Amr Alanwar
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17713v1 Announce Type: new Abstract: Agent evaluations and trace-based learning often compare outputs across transformed views through a post-response correspondence treated as neutral preprocessing. We show that this correspondence is a measurement intervention: omitting it can manufactu...

📖 Read original article


57. MAGPIE-Net: Predicting short-duration heavy-rainfall events in station neighborhoods from multitemporal FY-4A AGRI observations ​

Author: Xiang Lin, Yunying Li, Chengzhi Ye, Zitong Chen, Jing Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17753v1 Announce Type: new Abstract: Short-duration heavy-rainfall warning determines whether 1 h rainfall will exceed a threshold within a target-station neighborhood over the next few hours. Multitemporal infrared and water-vapor observations from the Fengyun-4A Advanced Geostationary R...

📖 Read original article


58. Training-Free Human-in-the-Loop Anomaly Detection via Memory Bank Correction ​

Author: Ayusha Abbas, Saram Abbas, Kabita Adhikari
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17775v1 Announce Type: new Abstract: Anomaly detectors are hardest to deploy exactly where training data is scarcest: a newly commissioned production line has a handful of verified "golden" samples and no machine-learning engineer on the factory floor. We present a training-free human-in-...

📖 Read original article


59. Debate Training Reduces Reward Hacking in RLAIF ​

Author: Zachary Kenton, Lili Janzer, Rory Greig, Tian Huey Teh, Kirill Tyshchuk, Jonah Brown-Cohen, Harri Edwards, Senthooran Rajamanoharan, Noah Y. Siegel, Natasha Jaques, Rohin Shah
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17776v1 Announce Type: new Abstract: We demonstrate that RL finetuning an LLM using debate, a two-player adversarial game between a generator and a critic adjudicated by a weaker LLM judge, reduces reward hacking compared to a reinforcement learning from AI feedback (RLAIF) baseline. Rewa...

📖 Read original article


60. Fourth-Moment Geometry of Rademacher Sums ​

Author: Peigan Gao, Jian Qian
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.PR

arXiv:2608.17802v1 Announce Type: new Abstract: Let $\varepsilon_1,\ldots,\varepsilon_n$ be independent Rademacher signs and let $a=(a_1,\ldots,a_n)\in\R^n$ satisfy the normalization below. For the normalized Rademacher sum, we determine how its higher moments depend on the fourth-order mass. Combin...

📖 Read original article


61. An Empirical Study of Reward Specification and Benchmark Reliability in GRPO-based LLM Unlearning ​

Author: Rub'en Balbastre, Juan Manuel Ordu~na, Mariano P'erez
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.17804v1 Announce Type: new Abstract: Practical LLM unlearning is usually evaluated through two objectives: suppress target-specific knowledge and preserve non-target utility. In generative QA, this leaves a third behavior underspecified: when a target-adjacent prompt admits a broader answ...

📖 Read original article


62. MotoSafety: Edge-AI with Learned Temporal Importance for Two-Wheeler Collision Risk Assessment Under Time Pressure ​

Author: Sumit S. Shevtekar, Chandresh K. Maurya, Gourab Sil, Subasish Das
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC

arXiv:2608.17823v2 Announce Type: new Abstract: Powered two-wheeler riders face critical safety challenges in low- and middle-income countries, yet limited studies exist on how cognitive stressors such as Time Pressure influence collision risk. We address this gap by introducing a comprehensive data...

📖 Read original article


63. Leveraging Association Context Retrieval in Knowledge Edit- ing to Build White-Box Attacks on LLMs ​

Author: Roman Maksimov, Vladimir Aletov, Vladimir Solodkin, Dmitry Bylinkin, Daniil Medyakov, Aleksandr Beznosikov
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17836v1 Announce Type: new Abstract: As large language models (LLMs) are granted increasing autonomy, it is essential to investigate methods that can induce unsafe behavior. We propose a novel white-box attack inspired by locate-then-edit approaches from the field of Knowledge Editing. Ou...

📖 Read original article


64. MoRAX: Mobility-based Representation Augmentation for Geospatial Foundation Models ​

Author: Ya Wen, Jixuan Cai, Yulun Zhou, Alec Kirkley
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.SI

arXiv:2608.17848v1 Announce Type: new Abstract: Geospatial Foundation Models (GFMs) are emerging as a powerful paradigm for learning semantically rich and geographically consistent visual and physical representations. However, their reliance on Earth-observation (EO) data leaves information about hu...

📖 Read original article


65. Efficient Resource Optimization for Split Federated Learning ​

Author: Wei Wei, Xianhao Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17849v1 Announce Type: new Abstract: Split federated learning (SFL) has emerged as a powerful paradigm for model training at the edge. However, SFL inherently involves discrete decision variables for model splitting and resource allocation, resulting in a challenging mixed-integer problem...

📖 Read original article


66. Dynamic Compression in Recurrent Networks ​

Author: Jyothish Pari, Ryan Bahlous-Boldi, Pulkit Agrawal
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17896v1 Announce Type: new Abstract: Recurrent models process long contexts efficiently by compressing their history into a fixed-size state, but modern architectures typically do so in a single causal pass over the sequence. Each input must therefore be compressed before the model knows ...

📖 Read original article


67. Hybrid ML for Lightweight Pre-Route Delay Estimation in Open-Source IC Design ​

Author: Marvin Castro Castro, Erick Carvajal Barboza
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17914v1 Announce Type: new Abstract: Static Timing Analysis (STA) is a critical step in the design flow of digital integrated circuits, however, obtaining accurate delay estimations can represent a challenge when limited information regarding physical design is available. In response, thi...

📖 Read original article


68. Efficient RLVR Scheduling via Graph-Structured Online Difficulty Estimation ​

Author: Zhizhao Liu, Zhiliang Tian, Xi Wang, Zhihua Wen, Yihang Xiong, Zhiquan Lai, Dongsheng Li
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.17941v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves the reasoning capabilities of large language models but relies on costly rollout exploration. Assigning the same exploration budget to samples with different difficulty levels is inefficien...

📖 Read original article


69. SIGMA: SHAP-Guided Implicit-Trajectory Generation for Metadata-Free LLM-Based AutoFE ​

Author: Xuan Zheng, Kento Uchida, Shinichi Shirakawa
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.17948v1 Announce Type: new Abstract: Recent research has leveraged Large Language Models (LLMs) to enhance Automated Feature Engineering (AutoFE) through semantic descriptions and trajectory-based prompting. However, there exist two challenges that limit their applicability and scalabilit...

📖 Read original article


70. An Omitted Mode Is a Rare Rule: The Sampling-Verification Danger Law in Continuous Code World Models ​

Author: Javier Aguilar Mart'in
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY

arXiv:2608.17956v1 Announce Type: new Abstract: In the Code World Model paradigm an LLM synthesizes an executable world model that a classical planner searches, and the model is accepted when it reproduces sampled transitions. We ask what that acceptance certifies in continuous control. We define th...

📖 Read original article


71. Understanding the Surprising Generalization Properties of Tabular Foundation Models ​

Author: Nour Shaheen, Junwei Ma, Alex Labach, Frank Hutter, Valentin Thomas, Anthony L. Caterini
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17957v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) increasingly rely on in-context learning, where a model receives labelled examples at inference time and predicts labels for new inputs without updating its weights. Existing TFMs are typically trained on either massive...

📖 Read original article


72. Too Sure to Be Safe: Model Calibration for Reliable Log Anomaly Detection ​

Author: Bin Li, Dongdong Wang, Siyang Lu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE

arXiv:2608.17965v1 Announce Type: new Abstract: Online log anomaly detection is critical for maintaining the reliability of large-scale computing systems. Although recent language model-based log anomaly detectors achieve strong detection performance, their confidence estimates remain poorly calibra...

📖 Read original article


73. Evaluating and improving crop-yield forecasting methods during extreme drought ​

Author: Shrey Gupta, Yi Ming, George Mohler
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17971v1 Announce Type: new Abstract: The impact of climate variability on food production has led to the creation of various forecasting models that uses machine learning (ML), numerical weather predictors (NWP) or a hybrid of ML-NWP models to identify structural and physical relationship...

📖 Read original article


74. Recirculation ​

Author: Michael C. Mozer, Shoaib Ahmed Siddiqui, Danny Sawyer, Sunny Sanyal, Rosanne Liu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.17981v1 Announce Type: new Abstract: We describe an inference-time architectural enhancement for off-the-shelf foundation models that markedly reduces perplexity and boosts accuracy across generation and reasoning tasks. Our approach incurs essentially no additional latency during generat...

📖 Read original article


75. Composing Flow-Matching Energies with Known Physics: Generation, OOD Detection, and Inversion on PDE Fields ​

Author: Yixuan Sun, Anirban Samaddar, Sandeep Madireddy
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph

arXiv:2608.18004v1 Announce Type: new Abstract: Probabilistic modeling of physical fields benefits from both a data-driven prior and known physical structure such as the governing equations. Energy-based models (EBMs) are a natural fit since energies compose additively, which enables augmenting phys...

📖 Read original article


76. Policy-Invariant Reward Shaping from LLM Feedback: A Framework for Hybrid RL Agents ​

Author: Christophe D. Hounwanou, John Emeka Eze, Ya'e U. Gaba
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.18008v1 Announce Type: new Abstract: Combining large language models with reinforcement learning is increasingly explored, yet the theoretical status of LLM-derived reward signals is often left implicit. We formalize the hybrid LLM-planner and RL-controller architecture as a Goal-Augmente...

📖 Read original article


77. Revisiting WEASEL 2.0: Reproduction, Sensitivity, and an Adaptive Ensemble-Size Rule ​

Author: Cian Higgins, Gerard Carrigan, Pinar Sungu Isiacik, Georgiana Ifrim
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.18021v1 Announce Type: new Abstract: WEASEL 2.0 is a dictionary-based time series classifier that combines dilated sliding windows with a randomised hyperparameter ensemble and a fixed-size dense feature representation. Two of its hyperparameter choices, the maximum ensemble size and the ...

📖 Read original article


78. Why GPT-Style Models Do Not Directly Transfer to Symbolic Music: Compression in the Wrong Coordinate System ​

Author: Yi Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SD

arXiv:2608.18025v1 Announce Type: new Abstract: GPT-style models achieve strong performance by representing language with finite vocabularies of reusable discrete tokens. This success has motivated symbolic music tokenizations to treat recurring musical structures, such as chords, motifs, and phrase...

📖 Read original article


79. TabNSM: Neural Sparse Mixer for Tabular Regression ​

Author: Ali Eslamian, Qiang Cheng
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CE

arXiv:2608.18026v1 Announce Type: new Abstract: Large-scale, high-dimensional tabular regression remains challenging: tree-based models are robust but lack end-to-end representation learning, while deep models enable flexible feature learning but often incur costly interaction modeling and sensitivi...

📖 Read original article


80. Optimize Your Sampling: Tuned Diffusion Sampling with Bayesian Optimization ​

Author: Travis Zhang, Christian Belardi, Justin Lovelace, Jin Peng Zhou, Saebyeol Shin, Carla P. Gomes, Kilian Q. Weinberger
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.18040v1 Announce Type: new Abstract: Sampling from a diffusion model typically requires many forward passes through a large neural network, making generation computationally expensive. While much work has focused on efficient solvers and samplers, comparatively little attention has been p...

📖 Read original article


81. The concentration game: Bayesian updating, regret, and information ​

Author: Akshay Balsubramani
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.GT, math.PR, math.ST, stat.TH

arXiv:2608.18061v1 Announce Type: new Abstract: We give a two-player zero-sum repeated game between a learner and nature whose value identity generates Bayesian updating and an exact accounting of exponential-weights regret at once, and supplies the comparator-class variational form that a wide clas...

📖 Read original article


82. Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs ​

Author: Christos Koutsiaris
Published: 8/20/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG

arXiv:2602.14784v1 Announce Type: cross Abstract: Breaking long documents into smaller segments is a fundamental challenge in information retrieval. Whether for search engines, question-answering systems, or retrieval-augmented generation (RAG), effective segmentation determines how well systems can...

📖 Read original article


83. Information Spreading in Diffusion Models from Effective Field Theory ​

Author: Navonil Neogi, Nabil Iqbal
Published: 8/20/2026, 4:00:00 AM
Categories: hep-th, cond-mat.stat-mech, cs.LG

arXiv:2608.14308v1 Announce Type: cross Abstract: We study score-matching diffusion models with a convolutional architecture. We argue that the inductive bias of locality means that the machinery of effective field theory from physics can be usefully applied to describe the denoising dynamics. We ap...

📖 Read original article


84. Advancing Health Equity through Multi-Level Fairness in Health Informatics ​

Author: Nick Souligne, Vignesh Subbian
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CY, cs.LG

arXiv:2608.16902v1 Announce Type: cross Abstract: The increasing integration of machine learning in healthcare has highlighted critical challenges related to fairness, transparency, and health equity. Specifically, the use of multi-level fairness techniques, which combine multiple bias mitigation st...

📖 Read original article


85. ComNetX: Local Hierarchical Adaptation for Dynamic Community Detection ​

Author: Aleksandr Konovalov, Anna Uporova, Alexander Drobyshev, Iaroslav Egorov, Grigoriy Bokov
Published: 8/20/2026, 4:00:00 AM
Categories: cs.SI, cs.AI, cs.LG

arXiv:2608.16906v1 Announce Type: cross Abstract: Dynamic community detection is commonly addressed either by full-snapshot recomputation or by solver-specific dynamic procedures. Full recomputation preserves the semantics of mature static solvers, but it repeatedly processes unchanged graph regions...

📖 Read original article


86. Which CS1 Students Will Fail? Identifying Digital Markers from Learning Analytics in Computer Systems and Architecture Using Weighted Academic Momentum and Interaction Logs ​

Author: Lighton Phiri, Mutune Chaibela, Ivy Chisha, David Pungwa, Danny Siabbaba, Bydon Simukoko
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CY, cs.LG

arXiv:2608.16914v1 Announce Type: cross Abstract: Digital learning platforms generate rich behavioural traces (digital markers) that offer the potential to identify struggling students early. This paper investigates whether a combination of traditional and digital markers can predict failure in a fi...

📖 Read original article


87. MITRE-SAGE: A Multi-Agent Cybersecurity Question-Answering Model ​

Author: Ali Habibzadeh, Farid Feyzi, Reza Ebrahimi Atani
Published: 8/20/2026, 4:00:00 AM
Categories: cs.IR, cs.CR, cs.LG, cs.MA

arXiv:2608.16921v2 Announce Type: cross Abstract: Effective cybersecurity operations require timely and accurate analysis of large-scale heterogeneous security information; however, analysts increasingly struggle with information overload, alert fatigue, and time-constrained decision-making. Althoug...

📖 Read original article


88. Network Denoising Revisited: A Ricci-Flow-Inspired Graph Diffusion Method ​

Author: Ye Fang, Chuan-Xian Ren
Published: 8/20/2026, 4:00:00 AM
Categories: cs.SI, cs.LG

arXiv:2608.16923v1 Announce Type: cross Abstract: Networks provide a fundamental representation of relationships among entities. However, real-world networks are often corrupted by noise caused by measurement errors and inherent stochasticity, hindering the discovery of meaningful structure. Most de...

📖 Read original article


89. SPSA Hyperparameter Tuning for Variational Quantum Natural Language Inference ​

Author: Nayan D'Souza, Christopher J. Agostino
Published: 8/20/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.16939v1 Announce Type: cross Abstract: Training variational quantum models requires choosing between parameter-shift gradients, which are exact but cost $O(P)$ forward evaluations, and simultaneous perturbation stochastic approximation (SPSA), which uses only two samples but produces high...

📖 Read original article


90. A Constant-Competitive Algorithm for Dynamic Mixture-of-Experts Serving ​

Author: Ian D'Ambrosio (Nth Research Collective)
Published: 8/20/2026, 4:00:00 AM
Categories: cs.DS, cs.LG

arXiv:2608.16947v1 Announce Type: cross Abstract: Huang, Lou, and Xiao introduced Dynamic Mixture-of-Experts Serving and gave an O(sqrt(log k))-competitive randomized algorithm for its integral primal problem, where k is the number of replica GPUs beyond the mandatory copy of each expert. Their matc...

📖 Read original article


91. WONDER: A Radio World Model-based Negotiation Framework for Multi-Agent UAV Coverage Optimization ​

Author: Jiahao Huang, Rongpeng Li, Zhifeng Zhao, Guoru Ding, Honggang Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.MA, cs.LG

arXiv:2608.16955v1 Announce Type: cross Abstract: Post-disaster damage to terrestrial infrastructure can disrupt wireless coverage,while Uncrewed Aerial Vehicle (UAV) swarms provide a promising solution for rapid restoration.However, due to the limitations in local geometry observations hidden radio...

📖 Read original article


92. The Price of Thinking: Reasoning Effort as a Model-Specific API Contract ​

Author: Yeabin Moon
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CY, cs.LG

arXiv:2608.16956v1 Announce Type: cross Abstract: API buyers purchase a dated contract, not a model name alone: the contract includes the requested and served model, reasoning-effort term or its omission, output rail, service product, prompt, and price schedule. We study the reasoning-effort term th...

📖 Read original article


93. MagViT: Interpretable Multi-Magnification Transformers with Patient-Level Model Selection for Breast Histopathology ​

Author: Nabil Ashab, Soumit Kumar Kundu, Saif Mahmud Parvez, Shahadat Hossain Sohag, Bidhan Biswas, Nazmus Subha
Published: 8/20/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2608.16959v1 Announce Type: cross Abstract: Breast cancer is one of the most common types of cancer among women around the world. Rapid detection and early treatment can hinder its progress to more complex stages and can impede its spread to other parts of the body. Histopathological image cla...

📖 Read original article


94. Diagonal Multi-omics Integration of Heterogenous Datasets ​

Author: Maksim V. Kukushkin, Mikhail S. Arbatskiy, Dmitriy E. Balandin, Alexey V. Churov
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.FA

arXiv:2608.16968v1 Announce Type: cross Abstract: In this paper, we consider methods for the diagonal multi-omics integration of heterogeneous datasets. Several approaches to the nature of biological heterogeneity are analyzed and developed to comprehend more clearly the generated differences. Speci...

📖 Read original article


95. Probing the Prefill: Detecting Code Vulnerabilities via Latent Activations ​

Author: Alizishaan Khatri
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.16970v1 Announce Type: cross Abstract: LLM-based code generation is now embedded in mission-critical pipelines, but defenses against vulnerable output remain post-hoc -- static analyzers, fine-tuned classifiers, or an LLM judge that screen completed code, ignoring the generating model's o...

📖 Read original article


96. FedPref: Federated Preference Learning for Structured Radiology Report Extraction ​

Author: Flint Xiaofeng Fan, Cheston Tan, Yew-Soon Ong, Roger Wattenhofer
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.16971v1 Announce Type: cross Abstract: Radiology reports describe findings and locations in free text, but downstream search and analysis require these relations in a fixed schema. Learning this extraction requires labels that are unevenly distributed across institutions: smaller hospital...

📖 Read original article


97. Margin-Regularized Structured Semantic Alignment for Brain-Language Correspondence ​

Author: Jiaqi Wang, Huawen Hu, Shu Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.16975v1 Announce Type: cross Abstract: With the rapid advancement of large language models, brain-language decoding has achieved remarkable progress. However, it remains unclear whether decoded content genuinely reflects neural representations or is largely reconstructed by the language m...

📖 Read original article


98. VLCP: Vision Language Control Policy Closed-Loop Code Replanning for Robot Manipulation ​

Author: Dhia Naouali, Minghan Wu, Claudia Wong, Abhinav Puthran, Omar G. Younis
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.16978v1 Announce Type: cross Abstract: Turning a frontier vision-language model into a robot policy usually means fine-tuning it to emit an action representation it never saw in pretraining, which throws away much of the reasoning that made the model worth reaching for. We go the other wa...

📖 Read original article


99. Lambda-Hold Control: Human-Like Movement Emerges from a Minimal Task Reward in Predictive Musculoskeletal Simulation ​

Author: Jun Hyuk Lee, Chihyeong Lee, Jooeun Ahn
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.GR, cs.LG

arXiv:2608.17030v1 Announce Type: cross Abstract: The massive overactuation in the human musculoskeletal system makes it challenging to train musculoskeletal models to generate human-like motion via reinforcement learning, primarily because exploration in the resulting high-dimensional and redundant...

📖 Read original article


100. Wasted large language models: A life cycle thinking approach ​

Author: Erik Johannes Husom, Maria Emine Nylund, Ophelia Prillard
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CY, cs.HC, cs.LG

arXiv:2608.17055v1 Announce Type: cross Abstract: Large Language Models (LLMs) are machine learning (ML) models that have an increasingly large carbon footprint through their development and use. Efforts to increase the energy efficiency of these models have not translated into reduced consumption d...

📖 Read original article


101. Dynamic Entanglement-Weighted Pruning for Quantum Federated Unlearning in Supply-Chain Risk Prediction ​

Author: Aditya Kumar, Sumit Chongder
Published: 8/20/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.17069v1 Announce Type: cross Abstract: Federated deployments of variational quantum classifiers are attractive for cross-organisation risk prediction in supply chains, because raw data never leaves the client, yet data-protection regulations such as the GDPR grant clients a right to reque...

📖 Read original article


102. Digital Twin-Based Intrusion Detection for Vehicle Powertrain CAN Bus Systems ​

Author: Araf Rahman, M Sabbir Salek, Mashrur Chowdhury
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.17093v1 Announce Type: cross Abstract: Existing automotive intrusion detection systems (IDSs) for the Controller Area Network (CAN) largely target discrepancies in message timing, frequency, or sequencing and cannot detect attacks that preserve these properties while manipulating the payl...

📖 Read original article


103. Picture the Epsilon: Pursuing Identity-Level Privacy Guarantees for Images ​

Author: Arman Zareian Jahromi, Vishnu Bondalakunta, Mohammad Akbar Bin Shah, Naimul Haque, Shuangqing Wei, George T. Amariucai
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.17147v1 Announce Type: cross Abstract: Image-to-image face generators are widely used, and visual dissimilarity between their outputs and source images is sometimes treated as evidence of privacy. Auditing whether these systems satisfy formal identity-level (epsilon, delta)-differential p...

📖 Read original article


104. Lymphocyte Mimicry Correction via Region-Level Tissue Reasoning and Unbalanced Optimal Transport ​

Author: Xiang Li, Yuqi Wang, Casey C. Heirman, Jihye Heo, Kyle J. Lafata
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.17151v1 Announce Type: cross Abstract: Cell mimicry arises when different cell types appear morphologically similar. Human pathologists resolve this ambiguity using surrounding tissue context, whereas current vision models either lack contextual reasoning (cell foundation models) or canno...

📖 Read original article


105. Policy Optimization and Statistical Inference for Online Contextual Matrix Games ​

Author: Liner Xiang, Yixin Wang, Hengrui Cai
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.ME, stat.TH

arXiv:2608.17173v1 Announce Type: cross Abstract: Online decision making often requires navigating a landscape shaped by both dynamic contexts and strategic interactions. In competitive pricing, for example, hotels must account for both dynamic contextual factors and rivals' strategic responses. Exi...

📖 Read original article


106. Expressivity In Multimodal Contrastive Learning ​

Author: Andrew Stuart, Florian Wolf
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2608.17203v1 Announce Type: cross Abstract: Contrastive learning has become a cornerstone of modern representation learning, powering CLIP-style models that underpin text-to-image generation, vision-language models, and retrieval across a rapidly growing range of modalities. Despite this empir...

📖 Read original article


107. Teach and Grow: An Agent-Centered Architecture for General Robot Learning ​

Author: Chang Nie, Zhe Liu, Hesheng Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG

arXiv:2608.17209v1 Announce Type: cross Abstract: End-to-end vision-language-action (VLA) and world-action models offer an elegant route to general-purpose robotics, but their reliability is bounded by validated physical coverage. When an unfamiliar object, sensor, embodiment, or contact falls outsi...

📖 Read original article


108. Temporal Leakage in Financial News NLP: A Multi-Architecture Audit with a Regime-Specific M&A Signal ​

Author: Chenhao Xue, Raslen Guesmi, Siwei Feng, Yucheng Gong, Jacob Xavier Sundram, Jordan Pang, Lan Wang, Julian Kaljuvee
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.17223v1 Announce Type: cross Abstract: Financial-news direction prediction has become a popular NLP benchmark, yet reported gains depend critically on whether the train-test split is chronological or random, i.e., on temporal leakage. We audit this dependence on a 49,799-article corpus ac...

📖 Read original article


109. Information fusion and machine learning for sensitivity analysis using physics knowledge and experimental data ​

Author: Berkcan Kapusuzoglu, Sankaran Mahadevan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CE, cs.LG, stat.ME, stat.ML

arXiv:2608.17248v1 Announce Type: cross Abstract: When computational models (either physics-based or data-driven) are used for the sensitivity analysis of engineering systems, the sensitivity estimate is affected by the accuracy and uncertainty of the model. This paper considers global sensitivity a...

📖 Read original article


110. Adaptive surrogate modeling for high-dimensional spatio-temporal output ​

Author: Berkcan Kapusuzoglu, Shunsaku Matsumoto, Yoshitomo Miyagi, Daigo Watanabe, Sankaran Mahadevan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CE, cs.AI, cs.LG, stat.ME, stat.ML

arXiv:2608.17250v1 Announce Type: cross Abstract: This paper develops an adaptive surrogate modeling method for problems with very high-dimensional spatio-temporal outputs. The analysis of spatio-temporal multi-physics systems is computationally expensive and consists of a large number of inputs and...

📖 Read original article


111. SPACE: Sample-cloud Predictive Adaptive Conformal Ellipsoids for Multivariate Time-Series Forecasting ​

Author: Baishi Li, Kelvin J. L. Koa, Ke-Wei Huang
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2608.17333v1 Announce Type: cross Abstract: Modern probabilistic time-series forecasters often express uncertainty through forecast samples. While typically converted into nominal prediction regions using empirical quantiles, these model-implied sets lack formal coverage guarantees and frequen...

📖 Read original article


112. On the Pseudo-Mixing of Kac's Walk ​

Author: Natesh S. Pillai, Aaron Smith, Vinod Vaikuntanathan
Published: 8/20/2026, 4:00:00 AM
Categories: math.PR, cs.CR, cs.LG

arXiv:2608.17374v1 Announce Type: cross Abstract: Motivated by a conjecture of Vaikuntanathan and Zamir, we study the pseudo-mixing of Kac's walk on $\mathrm{SO}(n)$: whether short trajectories are indistinguishable from Haar measure by low-complexity tests. We prove that the first $k$ columns mix i...

📖 Read original article


113. Leveraging generative hallucination and biophysics-informed modeling for unified biomolecular sequence-structure co-design ​

Author: Xuefeng Liu, Mingxuan Cao, Xiao Luo, Songhao Jiang, Tobin Sosnick, Jinbo Xu, Louis Maher, Rick Stevens
Published: 8/20/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG, q-bio.BM

arXiv:2608.17381v1 Announce Type: cross Abstract: Biomolecular design underpins applications from molecular recognition to therapeutics and synthetic biology, yet de novo interaction design remains challenging-especially for DNA/RNA, underexplored non-protein modalities with scarce, heterogeneous co...

📖 Read original article


114. Prism-GRPO: Faster VLA Policy Optimization via Splitting Same-outcome Groups ​

Author: Zeyun Deng, Yuzhe Lu, Yawei Wang, Linbo Liu, Qing Ping, Han Ding, Guande Wu, Panpan Xu, Jun Huan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.17423v1 Announce Type: cross Abstract: GRPO is increasingly used for reinforcement learning of vision-language-action (VLA) policies because, unlike PPO, it does not require training a critic. This simplification comes with a sampling cost: group-relative advantages require multiple rollo...

📖 Read original article


115. Nonlocal Transition Kernel for Efficient Learning of Restricted Boltzmann Machines ​

Author: Kaiji Sekimoto, Muneki Yasuda
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.dis-nn, cs.LG

arXiv:2608.17450v1 Announce Type: cross Abstract: Learning restricted Boltzmann machines (RBMs) is computationally challenging because it requires expectations whose exact evaluation is generally intractable. The expectations are typically evaluated using a sampling approximation based on blocked Gi...

📖 Read original article


116. Online Generalized Sparse Regression: How Does Overparametrization Help? ​

Author: Shuoguang Yang, Qiang Sun
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2608.17466v1 Announce Type: cross Abstract: Regularized sparse regression has been extensively studied in the offline setting, but online formulation remains relatively under-explored. This gap stems from four key challenges: (i) the infeasibility of dynamically updating the regularization par...

📖 Read original article


117. When AI Designs AI: Innovation or Imitation? ​

Author: Yikang Yang, Zhengxin Yang, Luzhou Peng, Minghao Luo, Yanqi Kan, Wanling Gao, Jianfeng Zhan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.17471v1 Announce Type: cross Abstract: Recent advances in LLM agents have made them increasingly capable of designing methods for complex AI tasks. This raises two central questions about agent-designed methods relative to human-designed methods: how well they perform, and how different t...

📖 Read original article


118. SGHA: Evidence-Grounded Research Problem Discovery with Local Language Models ​

Author: Sarvesh Gharat, Junpei Komiyama
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.17501v1 Announce Type: cross Abstract: Recent efforts toward fully automated AI scientists have demonstrated that language-model agents can generate hypotheses, execute experiments, and draft scientific manuscripts. However, during the early stages of research, when research problems are ...

📖 Read original article


119. Looking Beyond the Scale: Do Surgical Skill Models Learn Transferable Representations Across Assessment Rubrics? ​

Author: Hanna Hoffmann, Felix von Bechtolsheim, Stefanie Speidel, Rebecca Hisey
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.17519v1 Announce Type: cross Abstract: Vision-based surgical skill assessment has shown strong in-domain results, yet a fundamental question remains unasked: do these models learn transferable representations of surgical proficiency, or do they merely encode dataset-specific visual patter...

📖 Read original article


120. When to Review: Spaced Repetition for Continual Pre-Training of Language Models ​

Author: Alankar Atreya, Devesh Batra, Yoages Kumar Mantri, Geremy Bantug, Greig A Cowan, Raad Khraishi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.17530v1 Announce Type: cross Abstract: Continual pre-training of large language models must acquire new information without erasing old knowledge. Existing replay methods often choose a global old/new mixture and sample uniformly, ignoring that examples differ in how quickly they are forg...

📖 Read original article


121. Reflex-Guard: A Low-Latency Guardrail for LLM Prompt Safety Using Dense Semantic Embeddings ​

Author: Istiaque Ahmed, Afia Anjum Borsha, Ranat Das Prangon, Abu-fuad Ahmad, Thi Hong Tran
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG

arXiv:2608.17556v1 Announce Type: cross Abstract: Large Language Models (LLMs) in real-world applications often face the risks of specially crafted prompts designed to bypass the safety controls. Existing guardrail methods, such as LLM-as-a-judge and cloud-based safety APIs are able to detect unsafe...

📖 Read original article


122. Leveraging existing sparse point annotations for benthic imagery dense segmentation ​

Author: Cesar Borja, Breck A. McCollum, Jarret E. Byrnes, Kenneth Sebens, Ana C. Murillo
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.17561v1 Announce Type: cross Abstract: The health of marine ecosystems is a critical indicator of global environmental change, yet the physical constraints of underwater observation and the intrinsic challenges of processing marine imagery severely limit the scalability of systematic moni...

📖 Read original article


123. Feature Priming in Online Linear Regression: Sparse-Regret Lower Bounds and a Tight Univariate Rate ​

Author: Huibo Xu, Shi Fu, Qixin Zhang, Dacheng Tao
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP

arXiv:2608.17573v1 Announce Type: cross Abstract: In high-dimensional online prediction, the best predictor may depend on only a few features, so regret should scale with sparsity rather than the ambient dimension. Feature priming pursues this goal by estimating feature weights from past data and re...

📖 Read original article


124. Communication Reduction via Semantic-Based Encoding in DMPC Using LSTMs ​

Author: Torben Schiz, Pedro H. J. Nardelli, Henrik Ebel
Published: 8/20/2026, 4:00:00 AM
Categories: eess.SY, cs.DC, cs.LG, cs.MA, cs.RO, cs.SY

arXiv:2608.17592v1 Announce Type: cross Abstract: The communication demands of distributed model prediction control (DMPC) can overwhelm even advanced wireless communication technologies as agents must exchange a significant amount of information at least once per time step. To semantically reduce c...

📖 Read original article


125. MoNe: Modular Neural Memory for Efficient Long Context Inference ​

Author: Wonguk Cho, Kyubyung Chae, Tribhuvanesh Orekondy, Sunghyun Park, Hyoungwoo Park, Jeongho Kim, Arash Behboodi, Kyuwoong Hwang, Sungrack Yun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.17616v1 Announce Type: cross Abstract: We present MoNe, a lightweight modular neural memory that attaches to any frozen pretrained Transformer to enable long-context inference without retraining. MoNe reads context in fixed-size segments via test-time learning of fast-weight neural memory...

📖 Read original article


126. Iterative Grasp Pose Refinement: A Deep Reinforcement Learning Approach for 2D Vision ​

Author: Amir Arsalan Nematollahi, Shayan Ahmadi, Mehdi Tale Masouleh, Ahmad Kalhor
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, eess.IV

arXiv:2608.17628v1 Announce Type: cross Abstract: Developing robots capable of understanding and manipulating objects requires compact, interpretable, and generalizable representations. This work proposes a reinforcement learning-based framework for robotic grasp refinement, integrating keypoint-bas...

📖 Read original article


127. Mixture-of-Expert Blocks Contain Strong Hallucination Detection Signals ​

Author: Joao Fonseca, Rodrigo Rodrigues, Paolo Romano
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.17687v1 Announce Type: cross Abstract: Despite their widespread use, Large Language Models (LLMs) remain limited by a fundamental problem: the generation of plausible but false content, known as hallucinations. Most existing detection methods operate at the answer or sentence level, yet p...

📖 Read original article


128. MemCatalyst: Amplifying Data Auditing on Vision-Language Models via Data Poisoning ​

Author: Xukun Luan, Jinyan Liu, Yuhui Gong, Yuanguo Bi, Bing Hu, Xuesong Li, Di Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.17722v1 Announce Type: cross Abstract: Vision-Language models (VLMs) achieve outstanding performance largely due to the amount of training data available on the internet. At the same time, data holders (e.g., artists) urgently need to determine whether their data has been used for model t...

📖 Read original article


129. Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See ​

Author: Ayoub Kirouane, Christos Petrocheilos
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.RO, stat.ML

arXiv:2608.17744v1 Announce Type: cross Abstract: Take three frontier mixture-of-experts models (Alibaba, OpenAI, NVIDIA; 3.6-4.0B active parameters each) and fine-tune them to reason in a low-resource language. On accuracy benchmarks almost nothing happens, and the benchmark itself is noise at this...

📖 Read original article


130. Diff-DDoS: Realistic Cyber-Physical Attack Synthesis and Robust Detection for 5G-Enabled CPS Using Tabular Diffusion Models ​

Author: Bilal Hussain, Xiao Tang, Qinghe Du, Tan Li, Muhammad Azhar, Danista Khan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.17796v1 Announce Type: cross Abstract: Deep learning-based DDoS detectors for 5G-enabled cyber-physical systems face scarce labeled attack data and unrealistic synthetic substitutes, which limit robustness against adaptive adversaries. Detectors trained on hand-crafted attacks with fixed ...

📖 Read original article


131. Spatially explicit feature importance for building height estimation using research-access high-resolution SAR and optical sensors ​

Author: Guilherme Iablonovski, Pierre-Louis Frison, Tatiana Silva da Silva
Published: 8/20/2026, 4:00:00 AM
Categories: physics.soc-ph, cs.LG, stat.ML

arXiv:2608.17822v1 Announce Type: cross Abstract: Accurate building height information at the individual footprint scale is essential for material stock accounting and post-disaster damage assessments yet remains difficult to obtain at city scale in the Global South where airborne LiDAR coverage is ...

📖 Read original article


132. Toward the Optimal Regret-Instability Trade-off in Multi-Armed Bandits ​

Author: Kaifei Wang, Yinyu Ye, Han Zhong
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC, math.ST, stat.TH

arXiv:2608.17841v1 Announce Type: cross Abstract: Multi-armed bandit algorithms are evaluated by regret, yet comparable regret can coexist with different allocations across independent runs. We study the trade-off between worst-case regret $\mathcal{R}{K,T}$ and instability $\mathcal S$, defi...

📖 Read original article


133. A Residual Learning Approach for Unsteady Aerodynamic Load Prediction ​

Author: Divya Sanghi, Carlos E. S. Cesnik
Published: 8/20/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG

arXiv:2608.17894v1 Announce Type: cross Abstract: This paper investigates the feasibility of using residual learning to improve unsteady aerodynamic load prediction for aeroelastic applications. The machine learning technique selected for the study is the long short-term memory (LSTM) neural network...

📖 Read original article


134. AppendiGrade: An XAI-Enhanced Deep Learning Framework for Grading Appendicitis in Ultrasound with Gaussian Blur and Grad-CAM ​

Author: Fahad Ahammed, Omar Faruq Shikdar, Navid Zaman, Md Tahsin, Md. Nawab Yousuf Ali, Golam Sorwar
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.17923v1 Announce Type: cross Abstract: Appendicitis is one of the most common abdominal emergencies worldwide and requires prompt diagnosis and treatment to prevent life-threatening conditions. However, accurately differentiating complicated cases, such as perforation or abscess formation...

📖 Read original article


135. Procedural Content Metageneration via Program Search and Continual Abstraction Discovery ​

Author: Matthew Siper, Ahmed Khalifa, Julian Togelius
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NE

arXiv:2608.17947v1 Announce Type: cross Abstract: Large language models can generate executable programs, which makes it possible to search directly over procedural content generators rather than individual levels. We study this approach in Sokoban, Zelda, Dangerous Dave, and Lode Runner. Each run e...

📖 Read original article


136. Towards Zero-Shot Task Transfer with Neurosymbolic World Models ​

Author: Isidoro Tamassia, Lennert De Smet, Giuseppe Marra
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.17959v1 Announce Type: cross Abstract: State-of-the-art model-based reinforcement learning methods learn neural world models that allow policy improvement by planning in a latent space, without assumptions on the structure of the underlying environment. While expressive, these models are ...

📖 Read original article


137. Against Political Polarization: A Unified Framework for Tracing Evolving Political Ideologies on Social Media ​

Author: Yijie Xu, Chao Wang, Hui Xiong
Published: 8/20/2026, 4:00:00 AM
Categories: cs.SI, cs.AI, cs.CL, cs.LG

arXiv:2608.17987v1 Announce Type: cross Abstract: The rapid growth of social media has greatly influenced political discourse, highlighting the need to understand individual political ideologies and their temporal dynamics. This task faces challenges such as data scarcity, abundant non-political con...

📖 Read original article


138. Where A Small Language Model Helps in Invoice Categorisation, Understood Through Embedding Geometry ​

Author: Emma Ceccherini, Daniel Lawson, Anjulika Salhan
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.18033v1 Announce Type: cross Abstract: Categorising invoices into the correct General Ledger (GL) code underpins financial reporting and tax compliance. This is a skilled accounting judgement rather than a routine task: the correct category depends subtly on the nature of the purchasing b...

📖 Read original article


139. Harnessing Magnitude-Only and Complex Measurements for Improved Dynamic MRI Reconstruction with Learned Priors ​

Author: Mahdi Saberi, Ya\c{s}ar Utku Al\c{c}alar, Merve G"{u}lle, Chetan Shenoy, Mehmet Ak\c{c}akaya
Published: 8/20/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG, physics.med-ph

arXiv:2608.18036v1 Announce Type: cross Abstract: MRI reconstruction methods for undersampled k-space data naturally utilize complex-valued measurements. Parallel developments in sparse phase retrieval have shown that magnitude-only measurements may provide complementary information for signal recov...

📖 Read original article


140. Primitive Representation Learning for Unsupervised Dynamic Contrast Enhanced MRI Reconstruction ​

Author: Veronika Spieker, Wenqi Huang, Cemre Ariyurek, Liam Timms, Daniel Rueckert, Onur Afacan, Julia A. Schnabel, Sila Kurugol
Published: 8/20/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, eess.SP, physics.med-ph

arXiv:2608.18055v2 Announce Type: cross Abstract: Reliable quantitative analysis of dynamic contrast-enhanced MRI requires high-quality spatiotemporal reconstructions at high undersampling rates. Scan-specific reconstructions using Gaussian and Gabor primitives have shown promising results without t...

📖 Read original article


141. TokEval: A Tokenizer Evaluation Suite ​

Author: Clara Meister
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.18062v1 Announce Type: cross Abstract: Language model tokenizers are typically selected with minimal evaluation, despite the fact that their design choices directly impact model capabilities. This can be partly attributed to a limited understanding of which tokenizer properties affect whi...

📖 Read original article


142. On the Fragility of Self-Improving Agents: Variance, Task Order, and Underspecification ​

Author: Qinyuan Ye, Yu Li, Yada Pruksachatkun, Jiaxin Zhang, Chien-Sheng Wu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.18066v1 Announce Type: cross Abstract: Memory-based self-improving agents--those that learn from an online stream of tasks and improve over time by maintaining a textual memory bank--have shown great promise in recent literature. However, the reliability aspects of these methods have been...

📖 Read original article


143. HyPE-GT: where Graph Transformers meet Hyperbolic Positional Encodings ​

Author: Kushal Bose, Swagatam Das
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2312.06576v2 Announce Type: replace Abstract: Graph Transformers (GTs) facilitate the comprehension of complex relationships on graph-structured data by leveraging self-attention of the possible pairs of nodes. The structural information or inductive bias of the input graph is provided as posi...

📖 Read original article


144. Diffusion Models for Smarter UAVs: Decision-Making and Modeling ​

Author: Yousef Emami, Hao Zhou, Luis Almeida, Kai Li
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2501.05819v2 Announce Type: replace Abstract: Uncrewed Aerial Vehicles (UAVs) are increasingly used in modern communication networks. However, challenges in decision-making and digital modeling continue to hinder their rapid development. Reinforcement Learning (RL) algorithms face limitations ...

📖 Read original article


145. Gradient Heterogeneity Complements Hessian Heterogeneity in Transformer Optimization ​

Author: Akiyoshi Tomihari, Issei Sato
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE

arXiv:2502.00213v5 Announce Type: replace Abstract: Transformers are difficult to optimize with stochastic gradient descent (SGD) and largely rely on adaptive optimizers such as Adam. Despite extensive efforts, the mechanisms behind Adam's advantage over SGD in Transformer optimization are still not...

📖 Read original article


146. LZ Penalty: An information-theoretic repetition penalty for autoregressive language models ​

Author: Antonio A. Ginart, Naveen Kodali, Jason Lee, Caiming Xiong, Silvio Savarese, John R. Emmons
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT

arXiv:2504.20131v4 Announce Type: replace Abstract: We introduce the LZ penalty, a penalty specialized for reducing degenerate repetitions in autoregressive language models without loss of capability. The penalty is based on the codelengths in the LZ77 universal lossless compression algorithm. Throu...

📖 Read original article


147. TabularQGAN: A quantum generative model for tabular data synthesis ​

Author: Pallavi Bhardwaj, Caitlin Jones, Lasse Dierich, Aleksandar Vu\v{c}kovi'c
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, quant-ph

arXiv:2505.22533v2 Announce Type: replace Abstract: In this paper, we introduce a novel quantum generative model for synthesizing tabular data. Synthetic data is valuable in scenarios where real-world data is scarce or private, as it can be used to augment or replace existing datasets. As enterprise...

📖 Read original article


148. Global Convergence of Gradient EM for Over-Parameterized Gaussian Mixtures ​

Author: Mo Zhou, Weihang Xu, Maryam Fazel, Simon S. Du
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2506.06584v2 Announce Type: replace Abstract: Learning Gaussian Mixture Models (GMMs) is a fundamental problem in statistics and machine learning, with the Expectation-Maximization (EM) algorithm and its popular variant gradient EM being arguably the most widely used algorithms in practice. In...

📖 Read original article


149. Monotone Classification with Relative Approximations ​

Author: Yufei Tao
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2506.10775v3 Announce Type: replace Abstract: In monotone classification, the input is a multi-set $P$ of points in $\mathbb{R}^d$, each associated with a hidden label from ${-1, 1}$. The goal is to identify a monotone function $h$, which acts as a classifier, mapping from $\mathbb{R}^d$ to ...

📖 Read original article


150. Continuous Evolution Pool: Taming Recurring Concept Drift in Online Time Series Forecasting ​

Author: Tianxiang Zhan, Ming Jin, Yuanpeng He, Yuxuan Liang, Shirui Pan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2506.14790v3 Announce Type: replace Abstract: Recurring concept drift is pervasive in real-world online time series, where the underlying data-generating process repeatedly alternates between a small set of regimes, most notably daily or seasonal cycles that dominate energy, traffic, and weath...

📖 Read original article


151. Scientific Machine Learning of Chaotic Systems Learns Reduced-Order Equations for Neural Populations ​

Author: Anthony G. Chesebro, David Hofmann, Vaibhav Dixit, Earl K. Miller, Richard H. Granger, Alan Edelman, Christopher V. Rackauckas, Lilianne R. Mujica-Parodi, Helmut H. Strey
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, nlin.CD, q-bio.NC

arXiv:2507.03631v5 Announce Type: replace Abstract: Extracting interpretable mathematical models from complex dynamical systems is difficult, especially for chaotic dynamics observed with noisy experimental data. We present PEM-UDE, a method that combines prediction-error methodology with universal ...

📖 Read original article


152. Exact Reformulation and Optimization for Direct Metric Optimization in Binary Imbalanced Classification ​

Author: Le Peng, Yash Travadi, Chuan He, Ying Cui, Ju Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2507.15240v2 Announce Type: replace Abstract: For classification with imbalanced class frequencies, i.e., imbalanced classification (IC), standard accuracy is known to be misleading as a performance measure. While most existing methods for IC resort to optimizing balanced accuracy (i.e., the a...

📖 Read original article


153. Neural Operator-Based Nonlinear Nudging for Chaotic Dynamical Systems ​

Author: Jaemin Oh, Jinsil Lee, Youngjoon Hong
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2508.05778v2 Announce Type: replace Abstract: Nudging is an empirical data assimilation technique that incorporates an observation-driven control term into the model dynamics. The trajectory of the nudged system approaches the true system trajectory over time, even when the initial conditions ...

📖 Read original article


154. HeteRo-Select: Informativeness as the Participation Driver in Heterogeneous Federated Learning ​

Author: Md. Akmol Masud, Md Abrar Jahin, Mahmud Hasan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2508.06692v3 Announce Type: replace Abstract: Federated learning systems typically allocate gradient compression by link speed. This is sensible when bandwidth and data informativeness align. However, under non-IID data, these signals often decorrelate or invert. A bandwidth-driven allocator t...

📖 Read original article


155. Estimating Parameter Fields in Multi-Physics PDEs from Scarce Measurements ​

Author: Xuyang Li, Mahdi Masmoudi, Rami Gharbi, Nizar Lajnef, Vishnu Naresh Boddeti
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CE

arXiv:2509.00203v3 Announce Type: replace Abstract: Parameterized partial differential equations (PDEs) underpin the mathematical modeling of complex systems in diverse domains, including engineering, healthcare, and physics. A central challenge in using PDEs for real-world applications is to accura...

📖 Read original article


156. Asynchronous Message Passing for Addressing Oversquashing in Graph Neural Networks ​

Author: Kushal Bose, Swagatam Das
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.06777v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) suffer from oversquashing, where structural bottlenecks limit message propagation between distant nodes, hindering tasks that require long-range interactions. Existing remedies are limited: graph rewiring alters edge co...

📖 Read original article


157. One Pipeline, Many Transformers: Pattern-Specific Imputation Specialists for Tabular Missing Data ​

Author: Jacob Feitelberg, Dwaipayan Saha, Kyuseong Choi, Zaid Ahmad, Anish Agarwal, Raaz Dwivedi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.02625v5 Announce Type: replace Abstract: Missing data in tabular datasets forces practitioners into a hard choice: deploy a general-purpose imputer that may perform poorly for the problem at hand, or wait for someone to design a specialized algorithm. This problem is worsened by the fact ...

📖 Read original article


158. A Weak Penalty Neural ODE for Learning Chaotic Dynamics from Noisy Time Series ​

Author: Xuyang Li, John Harlim, Dibyajyoti Chakraborty, Romit Maulik
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.DS

arXiv:2511.06609v5 Announce Type: replace Abstract: The accurate forecasting of complex, high-dimensional dynamical systems from observational data is a fundamental task across numerous scientific and engineering disciplines. A significant challenge arises from noisy observations of deterministic dy...

📖 Read original article


159. Row-Stochastic Matrices Can Provably Outperform Doubly Stochastic Matrices in Decentralized Learning ​

Author: Bing Liu, Boao Kong, Limin Lu, Kun Yuan, Chengcheng Zhao
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.19513v4 Announce Type: replace Abstract: Decentralized learning often involves a weighted global loss with heterogeneous node weights $\lambda$. We revisit two natural strategies for incorporating these weights: (i) embedding them into the local losses to retain a uniform weight (and thus...

📖 Read original article


160. Cluster Aggregated GAN (CAG): A Cluster-Based Hybrid Model for Appliance Pattern Generation ​

Author: Zikun Guo, Adeyinka. P. Adedigba, Rammohan Mallipeddi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2512.22287v4 Announce Type: replace Abstract: Synthetic appliance data are essential for developing non-intrusive load monitoring algorithms and enabling privacy preserving energy research, yet the scarcity of labeled datasets remains a significant barrier. Recent GAN-based methods have demons...

📖 Read original article


161. Parametric and Generative Forecasts of EPEX Day-Ahead Energy Market Curves ​

Author: Julian Gutierrez, Redouane Silvente
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2601.20226v3 Announce Type: replace Abstract: We propose two methodologies for modelling aggregated supply and demand curves in the EPEX SPOT Day-Ahead market, emphasizing generative models as a way to recover distributional variability. The first is a low-dimensional parametric representation...

📖 Read original article


162. How (Not) to Hybridize Neural and Mechanistic Models for Epidemiological Forecasting ​

Author: Yiqi Su, Ray Lee, Jiaming Cui, Naren Ramakrishnan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.06323v3 Announce Type: replace Abstract: Epidemiological forecasting from surveillance data is a hard problem and hybridizing mechanistic compartmental models with neural models is a natural direction. The mechanistic structure helps keep trajectories epidemiologically plausible, while ne...

📖 Read original article


163. Community Concealment from Graph Neural Networks ​

Author: Dalyapraz Manatova, Pablo Moriano, L. Jean Camp
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.SI

arXiv:2602.12250v2 Announce Type: replace Abstract: Graph neural networks (GNNs) enable powerful unsupervised learning of communities. However, such inference may inadvertently expose sensitive group structures, critical clustered patterns, or collective behaviors, raising concerns about sensitive g...

📖 Read original article


164. TiMi: Empower Time Series Transformers with Multimodal Mixture of Experts ​

Author: Jiafeng Lin, Yuxuan Wang, Huakun Luo, Jianmin Wang, Zhongyi Pei
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.21693v2 Announce Type: replace Abstract: Multimodal time series forecasting has garnered significant attention for its potential to provide more accurate predictions than traditional single-modality models by leveraging rich information inherent in other modalities. However, due to fundam...

📖 Read original article


165. Quantifying Memorization and Privacy Risks in Genomic Language Models ​

Author: Alexander Nemecek, Wenbiao Li, Xiaoqian Jiang, Jaideep Vaidya, Erman Ayday
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, q-bio.GN

arXiv:2603.08913v2 Announce Type: replace Abstract: Genomic language models (GLMs) have emerged as powerful tools for learning representations of DNA sequences, enabling advances in variant prediction, regulatory element identification, and cross-task transfer learning. However, as these models are ...

📖 Read original article


166. How to make the most of your masked language model for protein engineering ​

Author: Calvin McCarter, Nick Bhattacharya, Sebastian W. Ober, Hunter Elliott
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM

arXiv:2603.10302v3 Announce Type: replace Abstract: A plethora of protein language models have been released in recent years. Yet comparatively little work has addressed how to best sample from them to optimize desired biological properties. We fill this gap by proposing a flexible, effective sampli...

📖 Read original article


167. Likelihood Hacking in Probabilistic Program Synthesis ​

Author: Jacek Karwowski, Younesse Kaddar, Zihuiwen Ye, Esmeralda S. Whitammer, Sam Staton
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.PL

arXiv:2603.24126v2 Announce Type: replace Abstract: When language models are trained by reinforcement learning (RL) to write probabilistic programs, they can artificially inflate their marginal-likelihood reward by producing programs whose data distribution fails to normalise instead of fitting the ...

📖 Read original article


168. Self-Distillation as a Performance Recovery Mechanism for LLMs: Counteracting Compression and Catastrophic Forgetting ​

Author: Chi Liu, Xin Chen, Xu Zhou, Fangbo Tu, Srinivasan Manoharan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2604.15794v2 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved remarkable success, underpinning diverse AI applications. However, they often suffer from performance degradation due to factors such as catastrophic forgetting during Supervised Fine-Tuning (SFT), quantiz...

📖 Read original article


169. Protect the Brain When Treating the Heart: Feasibility of 2.5D U-Net for Real-Time Gaseous Microemboli Detection ​

Author: Andrea Angino, Ken Trotti, Diego Ulisse Pizzagalli, Rolf Krause, Tiziano Torre, Stefanos Demertzis
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.22258v2 Announce Type: replace Abstract: Gaseous microemboli (GME) represent a common complication of cardiac structural interventions across both surgical and transcatheter approaches. Intraoperative transesophageal echocardiography (TEE) represents a convenient methodology to monitor an...

📖 Read original article


170. The Optimal Sample Complexity of Multiclass and List Learning ​

Author: Chirag Pabbaraju
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2604.24749v4 Announce Type: replace Abstract: While the optimal sample complexity of binary classification in terms of the VC dimension is well-established, determining the optimal sample complexity of multiclass classification has remained open. The appropriate complexity parameter for multic...

📖 Read original article


171. A Finite-Iteration Theory for Asynchronous Categorical Distributional Temporal-Difference Learning ​

Author: Ege C. Kaya, Abolfazl Hashemi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2605.06866v2 Announce Type: replace Abstract: We study finite-iteration behavior of the exact asynchronous recursions used by categorical distributional temporal-difference methods. The analysis covers scalar categorical TD in the Cram'er geometry and multivariate signed-categorical TD in the...

📖 Read original article


172. Latent Order Bandits ​

Author: Emil Carlsson, Newton Mwai, Fredrik D. Johansson
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.07304v2 Announce Type: replace Abstract: Bandit algorithms solve diverse sequential decision-making problems, but are often too sample-inefficient for from-scratch personalization. To substantially reduce exploration times, latent bandit algorithms exploit cross-instance structure implied...

📖 Read original article


173. Center-Manifold Reduction of Learning at Bifurcations: Interference and Rich Learning in Recurrent Neural Networks ​

Author: James Hazelden, Eric Shea-Brown
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.DS, math.OC, q-bio.NC

arXiv:2605.12763v2 Announce Type: replace Abstract: Rich learning in recurrent neural networks often proceeds through sudden transitions in latent dynamics, but there is little theory predicting how gradient descent behaves during these events. We study the local learning geometry near codimension-o...

📖 Read original article


174. FishBack: Pullback Fisher Geometry for Optimal Activation Steering in Transformers ​

Author: Sihan Wang, Jiayi Zhao, Qingyan Cao, Hongbo Yao, Lin Shu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2605.17231v2 Announce Type: replace Abstract: Activation steering has emerged as a lightweight approach for modifying language model behavior without parameter updates, yet existing methods remain brittle: unstable across layers and prone to disturbing behavior unrelated to the target concept....

📖 Read original article


175. ImplicitTerrainV2: Wavelet-Guided Spatially Adaptive Neural Terrain Representation ​

Author: Haoan Feng, Xin Xu, Leila De Floriani
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.22556v2 Announce Type: replace Abstract: Digital elevation models (DEMs) underpin terrain analysis in Geographic Information Systems (GIS), but commonly as raster representation, they rely on interpolation for off-grid sampling and finite-difference operators for derivative-based analysis...

📖 Read original article


176. Open datasets and machine learning for two-phase heat transfer: a review following a spatial-temporal taxonomy ​

Author: Christy Dunlap, Ridwan Olabiyi, Firas Al-Hindawi, Hari Pandey, Stephen Pierson, Daniel Curl, Braden Stevens, Mohammad Ishraq Hossain, Annapurna Parjuli, Chinmaya Joshi, Ashif Iquebal, Han Hu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn

arXiv:2605.23037v2 Announce Type: replace Abstract: Two-phase heat transfer underpins boiling, condensation, immersion cooling, flow boiling, energy conversion, and electronics thermal management, but its coupled interfacial physics make data reuse and model comparison difficult. This narrative revi...

📖 Read original article


177. Nonlinear Data Integration via Kernel Methods for Data Collaboration Analysis ​

Author: Yamato Suetake, Yuta Kawakami, Shunnosuke Ikeda, Yuichi Takano
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2605.27219v2 Announce Type: replace Abstract: Collaborative analysis of decentralized confidential datasets is important, but direct sharing of original datasets is often restricted by privacy and institutional constraints. Data collaboration (DC) analysis transforms each dataset into privacy-...

📖 Read original article


178. TabCausal: Pretraining Across Causal Environments for Tabular Causal Discovery ​

Author: Zi-Rong Li, Si-Yang Liu, Tian-Zuo Wang, Han-Jia Ye
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.31156v2 Announce Type: replace Abstract: Causal discovery aims to recover directed causal relations from observational and interventional data, providing a basis for mechanistic understanding and reliable decision-making. Causal discovery foundation models (CDFMs) seek to amortize this pr...

📖 Read original article


179. BRo-JEPA: Learning Modular Transformations in Latent Space ​

Author: Divyansh Jha, Yuanfang Xie, Brennen Yu, Varan Mehra
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2606.01372v2 Announce Type: replace Abstract: Can neural networks learn algebraic rules from visual inputs, or do they merely fit observed patterns? We study this question using MNIST (or EMNIST letters) as states and modular arithmetic operations as actions in a JEPA-style world model. Standa...

📖 Read original article


180. Mos-Gen: A Generative Molecular Framework for Mosquito Insecticide Design ​

Author: Lina Wang, Yaning Cui, Zhifeng Gao, Ping Xing, Biao Jiang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.01846v2 Announce Type: replace Abstract: Mosquito-borne infectious diseases cause more than 700000 deaths worldwide each year. The long-term use of conventional chemical insecticides has induced serious resistance problems, creating an urgent need to develop novel, highly effective, and e...

📖 Read original article


181. The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics ​

Author: Pietro Barbiero, Giovanni De Felice, Mateo Espinosa Zarlenga, Francesco Giannini, Filippo Bonchi, Mateja Jamnik, Giuseppe Marra, Ruggero Noris
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE

arXiv:2606.12289v2 Announce Type: replace Abstract: As Artificial Intelligence models grow in complexity, interpretability has become an indispensable tool for understanding, debugging, and controlling their computations. However, interpretability lacks general theories to deductively design interpr...

📖 Read original article


182. How Transparent is DiffusionGemma? ​

Author: Joshua Engels, Callum McDougall, Bilal Chughtai, Janos Kramar, Senthooran Rajamanoharan, Cindy Wu, Arthur Conmy, Asic Q Chen, Jean Tarbouriech, Min Ma, Brendan O'Donoghue, Jo~ao Gabriel Lopes de Oliveira, Rohin Shah, Neel Nanda
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.20560v2 Announce Type: replace Abstract: LLM reasoning transparency is a critical affordance for understanding model decisions, mitigating misuse and misalignment, and debugging surprising model behaviors. However, DiffusionGemma performs a larger fraction of its computation in a continuo...

📖 Read original article


183. Anti-Collapse Dynamics and the Emergence of Multi-Time-Scale Learning in Recurrent Neural Networks ​

Author: Lorenzo Livi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, physics.data-an

arXiv:2606.29519v3 Announce Type: replace Abstract: Long-range learning is hard for recurrent networks trained with stochastic gradient descent, because the influence of a past input fades with the lag $\ell$, and if it fades too fast the dependence cannot be learned from finite data. This fade is c...

📖 Read original article


184. Low-dimensional topology of deep neural networks ​

Author: Junyu Ren, Lek-Heng Lim
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.GT

arXiv:2606.31856v2 Announce Type: replace Abstract: We study layered models, including feedforward networks, ResNets, and transformers, by limiting each layer to a width of $d = 3$, i.e., $\mathbb{R}^3$ as representation space. This allows us to track how a neural network changes low-dimensional top...

📖 Read original article


185. SEE: Structure-aware Exploring & Exploiting for Long-horizon GUI Agent Trajectory Synthesis ​

Author: Zhuohang Fan, Beichen Zhang, Yuanfa Li, Changqiao Wu, Wei Liu, Jian Luan, Weigang Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.18046v2 Announce Type: replace Abstract: Graphical User Interface (GUI) agents powered by vision-language models hold promise for automating real-world mobile tasks. However, progress is limited by the lack of high-coverage, long-horizon interaction trajectories collected from element-ric...

📖 Read original article


186. GEqTrain: A Configuration-Driven Framework for Retargeting Equivariant Graph Neural Networks Across 3D Scientific Tasks ​

Author: Daniele Angioletti, Marco Nobile, Vittorio Limongelli
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM

arXiv:2607.19083v3 Announce Type: replace Abstract: Equivariant graph neural networks provide a powerful modeling language for three-dimensional scientific data, but their reuse is often limited by implementations tied to specific tasks, outputs, and training regimes. We present GEqTrain, a configur...

📖 Read original article


187. ClockRoPE: Random Fourier Rotations for Temporal Routine Modeling ​

Author: Yiwen Chen, Joshua Ainslie, Krzysztof Choromanski, Xiang Gao, Su-Lin Wu, Yiping Yuan, Qian Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.26369v2 Announce Type: replace Abstract: Rotary Position Embedding (RoPE) has been widely adopted in transformer-based large language models. However, its log-linear frequency schedule, originally designed to produce long-term attention decay, limits its adoption in domains with more comp...

📖 Read original article


188. A Physics-Informed Hybrid Neural Operator for Transient Magnetization Prediction in Power Magnetics ​

Author: Yachao Zhu, Qiujie Huang, Sinan Li, Yang Li, Gang Lei, Jianguo Zhu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.SY, eess.SY

arXiv:2608.02965v3 Announce Type: replace Abstract: Magnetic components in high-frequency, high-power-density converters are increasingly driven by non-sinusoidal flux-density waveforms with fast transitions, minor-loop operation, dc bias, and temperature variation. Under these conditions, steady-st...

📖 Read original article


189. A Comparative Study of Feature Selection Methods for EHR Diagnosis Codes in Opioid Use Disorder Prediction ​

Author: Zihan Ding, Yinan Liu, Tengfei Ma, Rachel Wong, Xia Zhao, Richard N. Rosenthal, Fusheng Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04180v2 Announce Type: replace Abstract: Feature selection is a critical step in electronic health record (EHR)-based predictive modeling, where input variables are often high-dimensional, sparse, noisy, and redundant. Large feature sets not only increase computational burden and overfitt...

📖 Read original article


190. Online Learning of Scale Parameters in Score-Driven Filters ​

Author: Fabrizio Lillo, Giulia Livieri, Gianluca Palmari
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ME, stat.ML, stat.TH

arXiv:2608.09218v2 Announce Type: replace Abstract: Score-driven filters update a time-varying parameter by multiplying a scaled log-likelihood score by a scale parameter that controls the magnitude of the update. We name this scale parameter gain, consider it a decision variable, and study its onli...

📖 Read original article


191. Geometric and Behavioral Stratification in Transformer Residual Streams ​

Author: Nelson Guda
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.12447v2 Announce Type: replace Abstract: Trained transformer models develop privileged bases: coordinate axes whose statistics differ from the rest of the residual stream. But what kind of direction does such a basis select? We investigate the prediction direction, the unembedding directi...

📖 Read original article


192. A Contract-Grade Verifier for LLM-Generated GPU Kernels, and a Native Blackwell Backward for the Gated-Linear-Recurrence Family ​

Author: Rishi Shah, Rishav Shrestha
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.DC

arXiv:2608.12700v2 Announce Type: replace Abstract: Systems that generate GPU kernels with language models report high correctness rates. Those rates come from a single loose test: run the kernel on a few random inputs at one fixed shape and accept it if the output is close to a reference. A kernel ...

📖 Read original article


193. Federated Compositional Muon Optimizer for Matrix-Wise Models ​

Author: Wang Yan, Feihu Huang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.12710v2 Announce Type: replace Abstract: Muon, a more recently developed optimizer, is useful for matrix-wise models in AI areas. Although many works have studied Muon and its variants, these methods are still not particularly well-suited for hierarchical structured problems. To fill this...

📖 Read original article


194. CoMedBench: A Multi-Source Benchmark of Synthetic Medical Data Fidelity and Downstream Utility ​

Author: Akanta Das, Farhad Al-Amin Dipto, Mrinmoy Sarkar Anto, David Rehkopf, Ayin Vala, Tanmoy Sarkar Pias
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12805v2 Announce Type: replace Abstract: Access to clinical data is essential for developing reliable healthcare machine learning systems, but direct use of electronic health records is constrained by privacy regulation, institutional review, data-use agreements, and the risk of re-identi...

📖 Read original article


195. The Integer Alibi: Localizing Cross-Kernel Divergence in INT8-Quantized LLM Inference ​

Author: Teng-Ruei Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13756v2 Announce Type: replace Abstract: Two GPU kernels implementing the same scaled INT8 GEMM interface are usually treated as interchangeable. We test that assumption: holding the checkpoint, prompts, hardware, inference engine, decoding, and quantization configuration fixed, we swap o...

📖 Read original article


196. CrevasseSeg: A Label-Efficient UAV Crevasse Segmentation Framework ​

Author: Steven Wallace, William D. Harcourt, Richard Hann, Aiden Durrant, Somayajulu Sripada, Georgios Leontidis
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15790v2 Announce Type: replace Abstract: Crevasse mapping from uncrewed aerial vehicle (UAV) imagery matters for glaciological research and for field safety in glaciated terrain. Yet, pixel-level annotation of glacier surfaces is costly and requires domain experts. We introduce CrevasseSe...

📖 Read original article


197. NICE: Scale-Stable Perturbations for Graph Neural Network Explanations via Noise Corruption ​

Author: Ziluowen Luo, Jun Yin, Ruochen Liu, Ming Cheng, Shirui Pan, Chengqi Zhang, Senzhang Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16038v2 Announce Type: replace Abstract: Post-hoc Graph Neural Network (GNN) explainers commonly follow a Perturb-Query paradigm, inferring the importance of graph elements based on queried predictions to perturbed inputs. However, such perturbations often introduce substantial distributi...

📖 Read original article


198. OceanDepths: A Global Dataset of Paired Subsurface and Surface Ocean Observations ​

Author: Simon Donike, Ruben Cartuyvels, Antonino Ian Ferola, Elisa Carli, Diego Fernandez Prieto, Marie-Helene Rio
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.16373v2 Announce Type: replace Abstract: Despite comprising over 70% of its surface, the world's oceans are critically underobserved compared to the land surface or the atmosphere. Understanding the global ocean requires jointly observing its surface and subsurface structure, yet no stand...

📖 Read original article


199. A Data-Efficient Analytical Prior Machine Learning Framework for Sound Reduction Frequency Prediction in Helmholtz Resonators ​

Author: Jiaming Li
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16873v2 Announce Type: replace Abstract: High-fidelity finite-element simulations can provide accurate numerical predictions for side-branch resonators, but large simulation datasets are expensive to generate and purely data-driven surrogates may become unreliable when simulation-labelled...

📖 Read original article


200. Deep Learning Based on Generative Adversarial and Convolutional Neural Networks for Financial Time Series Predictions ​

Author: Wilfredo Tovar
Published: 8/20/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2008.08041v3 Announce Type: replace-cross Abstract: In the big data era, deep learning and intelligent data mining technique solutions have been applied by researchers in various areas. Forecast and analysis of stock market data have represented an essential role in today's economy, and a sign...

📖 Read original article


201. The Authenticity Gap in Human Evaluation ​

Author: Kawin Ethayarajh, Dan Jurafsky
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2205.11930v3 Announce Type: replace-cross Abstract: Human ratings are the gold standard in NLG evaluation. The standard protocol is to collect ratings of generated text, average across annotators, and rank NLG systems by their average scores. However, little consideration has been given as to ...

📖 Read original article


202. Doubly robust nearest neighbors in factor models ​

Author: Raaz Dwivedi, Caleb Chin, Sabina Tomkins, Predrag Klasnja, Susan Murphy, Devavrat Shah
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2211.14297v5 Announce Type: replace-cross Abstract: We introduce and analyze an improved variant of nearest neighbors (NN) for estimation with missing data in latent factor models. We consider a matrix completion problem with missing data, where the $(i, t)$-th entry, when observed, is given b...

📖 Read original article


203. EquiPocket: an E(3)-Equivariant Geometric Graph Neural Network for Ligand Binding Site Prediction ​

Author: Yang Zhang, Zhewei Wei, Ye Yuan, Chongxuan Li, Wenbing Huang
Published: 8/20/2026, 4:00:00 AM
Categories: q-bio.BM, cs.LG

arXiv:2302.12177v5 Announce Type: replace-cross Abstract: Predicting the binding sites of target proteins plays a fundamental role in drug discovery. Most existing deep-learning methods consider a protein as a 3D image by spatially clustering its atoms into voxels and then feed the voxelized protein...

📖 Read original article


204. Comprehensive framework for evaluation of deep neural networks in detection and quantification of lymphoma from PET/CT images: clinical insights, pitfalls, and observer agreement analyses ​

Author: Shadab Ahamed, Yixi Xu, Sara Kurkowska, Claire Gowdy, Joo H. O, Ingrid Bloise, Don Wilson, Patrick Martineau, Fran\c{c}ois B'enard, Fereshteh Yousefirizi, Rahul Dodhia, Juan M. Lavista, William B. Weeks, Carlos F. Uribe, Arman Rahmim
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2311.09614v5 Announce Type: replace-cross Abstract: This study addresses critical gaps in automated lymphoma segmentation from PET/CT images, focusing on issues often overlooked in existing literature. While deep learning has been applied for lymphoma lesion segmentation, few studies incorpora...

📖 Read original article


205. Spikformer V2: Join the High Accuracy Club on ImageNet with an SNN Ticket ​

Author: Zhaokun Zhou, Yijie Lu, Kaiwei Che, Wei Fang, Keyu Tian, Qihao Peng, Yuesheng Zhu, Shuicheng Yan, Yonghong Tian, Li Yuan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.NE, cs.CV, cs.LG

arXiv:2401.02020v2 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs), known for their biologically plausible architecture, face the challenge of limited performance. The self-attention mechanism, which is the cornerstone of the high-performance Transformer and also a biologically...

📖 Read original article


206. Predicting Male Domestic Violence Using Explainable Ensemble Learning and Exploratory Data Analysis ​

Author: Md Abrar Jahin, Saleh Akram Naife, Fatema Tuj Johora Lima, M. F. Mridha, Md. Jakir Hossen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CY, cs.LG

arXiv:2403.15594v4 Announce Type: replace-cross Abstract: Domestic violence is commonly viewed as a gendered issue that primarily affects women, which tends to leave male victims largely overlooked. This study presents a novel, data-driven analysis of male domestic violence (MDV) in Bangladesh, high...

📖 Read original article


207. On Stability in Optimistic Bilevel Optimization ​

Author: Johannes O. Royset
Published: 8/20/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.SY, eess.SY

arXiv:2408.13323v3 Announce Type: replace-cross Abstract: Solutions of bilevel optimization problems tend to suffer from instability under changes to problem data. In the optimistic setting, we construct a lifted formulation that exhibits desirable stability properties under mild assumptions that ne...

📖 Read original article


208. Efficient Dynamic Shielding for Parametric Safety Specifications ​

Author: Davide Corsi, Kaushik Mallik, Andoni Rodriguez, Cesar Sanchez
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO, cs.RO, cs.SY, eess.SY

arXiv:2505.22104v2 Announce Type: replace-cross Abstract: Shielding has emerged as a promising approach for ensuring safety of AI-controlled autonomous systems. The algorithmic goal is to compute a shield, which is a runtime safety enforcement tool that needs to monitor and intervene the AI controll...

📖 Read original article


Author: Tuomas Kelom"aki, Abel Lacabanne, Daniel Tubbenhauer, Pedro Vaz, Victor L. Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: math.GT, cs.LG, math.QA

arXiv:2509.05574v3 Announce Type: replace-cross Abstract: We prove that, for many standard link invariants, both the proportion of distinct invariant values and the detection probability among prime alternating links with at most n crossings decay exponentially in n, with an explicit universal rate....

📖 Read original article


210. SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA ​

Author: Haozhou Xu, Dongxia Wu, Matteo Chinazzi, Ruijia Niu, Rose Yu, Yi-An Ma
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2509.25459v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) show promise in generating long-form scientific explanations that synthesize evidence and connect multiple factors. However, in long-form scientific question answering, LLMs often hallucinate, producing unsupporte...

📖 Read original article


211. Inverse Problems for Partial Differential Equations with Jump Discontinuities in Coefficients via Two-Stage Physics-Informed Deep Learning and Statistical Mixture Models ​

Author: Zhikun Zhang, Guanyu Pan, Xiangjun Wang, Yong Xu, Guangtao Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2510.14656v3 Announce Type: replace-cross Abstract: This work proposes a two-stage physics-informed deep learning framework that combines neural-network-based sampling with statistical inference and constrained parameter refinement. In the first stage, a dual-network physics-informed architect...

📖 Read original article


212. A multi-view contrastive learning framework for spatial embeddings in risk modelling ​

Author: Freek Holvoet, Christopher Blier-Wong, Katrien Antonio
Published: 8/20/2026, 4:00:00 AM
Categories: q-fin.RM, cs.LG

arXiv:2511.17954v2 Announce Type: replace-cross Abstract: Incorporating spatial information, particularly when related to climate, weather, and demographic factors, is crucial for improving underwriting precision and enhancing risk management in insurance. However, spatial data are often unstructure...

📖 Read original article


213. SparsePixels: Efficient Convolution for Sparse Data on FPGAs ​

Author: Ho Fung Tsoi, Dylan Rankin, Vladimir Loncar, Philip Harris
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, hep-ex

arXiv:2512.06208v4 Announce Type: replace-cross Abstract: Inference of standard convolutional neural networks (CNNs) on FPGAs often incurs high latency and a long initiation interval due to the deep nested loops required to densely convolve every input pixel regardless of its feature value. However,...

📖 Read original article


214. Large Language Models: A Mathematical Formulation ​

Author: Ricardo Baptista, Andrew Stuart, Son Tran
Published: 8/20/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, stat.ML

arXiv:2601.22170v2 Announce Type: replace-cross Abstract: Large language models (LLMs) process and predict sequences containing text to answer questions, and address tasks including document summarization, providing recommendations, writing software and solving quantitative problems. We provide a ma...

📖 Read original article


215. Fermi-Dirac thermal measurements: A framework for quantum hypothesis testing and semidefinite optimization ​

Author: Nana Liu, Mark M. Wilde
Published: 8/20/2026, 4:00:00 AM
Categories: quant-ph, cond-mat.stat-mech, cs.LG

arXiv:2603.04061v2 Announce Type: replace-cross Abstract: Quantum measurements are the means by which we recover messages encoded into quantum states. They are at the forefront of quantum hypothesis testing, wherein the goal is to perform an optimal measurement for arriving at a correct conclusion. ...

📖 Read original article


216. Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs? ​

Author: Jeonghye Kim, Xufang Luo, Minbeom Kim, Sangmook Lee, Dohyung Kim, Jiwon Jeon, Dongsheng Li, Yuqing Yang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2603.24472v4 Announce Type: replace-cross Abstract: Self-distillation has emerged as an effective post-training paradigm for LLMs, often improving performance while shortening reasoning traces. However, in mathematical reasoning, we find that it can reduce response length while degrading perfo...

📖 Read original article


217. From Diffusion to Flow: Efficient Motion Generation in MotionGPT3 ​

Author: Jaymin Bhan, JiHong Jeon, SangYeop Jeong
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2603.26747v3 Announce Type: replace-cross Abstract: Recent text-driven motion generation methods span both discrete token-based approaches and continuous-latent formulations. MotionGPT3 exemplifies the latter paradigm, combining a learned continuous motion latent space with a diffusion-based p...

📖 Read original article


218. Does Unification Come at a Cost? Uni-SafeBench: A Safety Benchmark for Unified Multimodal Large Models ​

Author: Zixiang Peng, Yongxiu Xu, Qin-Yi Zhang, Jiexun Shen, Yi-Fan Zhang, Hongbo Xu, Yubin Wang, Gaopeng Gou
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2604.00547v2 Announce Type: replace-cross Abstract: Unified Multimodal Large Models (UMLMs) integrate understanding and generation capabilities within a single architecture. While unified architectures expand multimodal capabilities, their safety implications remain important yet underexplored...

📖 Read original article


219. Attention Flows: Tracing LLM Conceptual Engagement via Story Summaries ​

Author: Rebecca M. M. Hicke, Sil Hamilton, David Mimno, Ross Deans Kristensen-McLachlan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.06416v2 Announce Type: replace-cross Abstract: Although LLM context lengths have grown, there is evidence that their ability to integrate information across long-form texts has not kept pace. We evaluate one such understanding task: generating summaries of novels. When human authors of su...

📖 Read original article


220. SegWithU: Uncertainty as Perturbation Energy for Single-Forward-Pass Risk-Aware Medical Image Segmentation ​

Author: Tianhao Fu, Austin Wang, Charles Chen, Roby Aldave-Garza, Yucheng Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2604.15271v4 Announce Type: replace-cross Abstract: Reliable uncertainty estimation is critical for medical image segmentation, where automated contours feed downstream quantification and clinical decision support. Many strong uncertainty methods require repeated inference, while efficient sin...

📖 Read original article


221. FairNVT: Fair Classification via Noise Injection in Vision Transformers ​

Author: Qiaoyue Tang, Sepidehsadat Hosseini, Mengyao Zhai, Thibaut Durand, Greg Mori
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2604.16780v2 Announce Type: replace-cross Abstract: This paper presents FairNVT, a lightweight debiasing framework for pretrained transformer-based encoders that improves prediction fairness while preserving task performance. FairNVT is motivated by the intuition that reducing sensitive-attrib...

📖 Read original article


222. Convergent Evolution: How Different Language Models Learn Similar Number Representations ​

Author: Deqing Fu, Tianyi Zhou, Mikhail Belkin, Vatsal Sharan, Robin Jia
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.20817v2 Announce Type: replace-cross Abstract: Language models trained on natural text learn to represent numbers using periodic features with dominant periods at $T=2, 5, 10$. In this paper, we identify a two-tiered hierarchy of these features: while Transformers, Linear RNNs, LSTMs, and...

📖 Read original article


223. Nonlinear GENERIC-Embedded Neural Networks (N-GENNs): Learning GENERIC dynamics with non-quadratic dissipation potentials ​

Author: Vojt\v{e}ch Votruba, Zequn He, Weilun Qiu, Celia Reina, Michal Pavelka
Published: 8/20/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG

arXiv:2605.09058v2 Announce Type: replace-cross Abstract: We introduce Nonlinear GENERIC-Embedded Neural Networks (N-GENNs), a deep learning framework for discovering evolution equations of systems governed by the nonlinear GENERIC formalism (General Equation for Non-Equilibrium Reversible-Irreversi...

📖 Read original article


224. Adaptive AI Task Partitioning and Safe Offloading in Heterogeneous Edge-Cloud Continuum ​

Author: Akuen Akoi Deng, Eimantas Butkus, Alfreds Lapkovskis, Praveen Kumar Donta
Published: 8/20/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG, cs.NI, cs.PF

arXiv:2605.09623v2 Announce Type: replace-cross Abstract: In recent years, the use of artificial intelligence on resource-constrained IoT devices has grown significantly. However, existing approaches to AI task partitioning and offloading across the edge-cloud continuum typically rely on static meth...

📖 Read original article


225. Memory by Design: Probabilistic Sequence Layers ​

Author: Matthew Dowling, Hyungju Jeon, Cristina Savin, Il Memming Park
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2605.31163v3 Announce Type: replace-cross Abstract: We introduce the \emph{design-model framework}: a way to derive efficient recurrent sequence maps from explicit assumptions about memory. A design model writes evidence into memory by exact Bayesian filtering; a query- dependent readout produ...

📖 Read original article


226. SEAM: Shortcut-Aware Real-Time Detection of Scripted vs. Spontaneous Speech for Interview Guardrails ​

Author: Vsevolod (V.), Kovalev, Pranay Manocha
Published: 8/20/2026, 4:00:00 AM
Categories: eess.AS, cs.LG

arXiv:2606.06837v2 Announce Type: replace-cross Abstract: Scripted vs spontaneous speech detection is appealing for interview guardrails, but benchmark performance can be inflated by shortcuts tied to corpus identity, channel conditions, and recording artifacts rather than speaking style itself. We ...

📖 Read original article


227. Spectrally Safe Neural Operator Warm-Starts for Large-Scale Newton Solvers ​

Author: Jaemin Oh, Youngkyu Lee, Jerome Darbon, George Em Karniadakis
Published: 8/20/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA

arXiv:2606.21828v2 Announce Type: replace-cross Abstract: Neural operators are increasingly used to warm-start Newton solvers for nonlinear PDEs, on the premise that a low test error places the initial guess inside the basin of attraction. We show that this premise is unreliable. An operator trained...

📖 Read original article


228. Turning Off-Policy Tokens On-Policy: A Plug-in Approach for Improving LLM Alignment ​

Author: Yu Li, Xiuyu Li, Mingyang Yi, Jiaxing Wang, Liangxu Zhang, Zhaolong Xing, Zhen Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.04728v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) post-training for large language models (LLMs) follows a efficient paradigm of "rollout then update", which inevitably results in off-policy training data. To resolve this, Importance sampling (IS) is proposed, whi...

📖 Read original article


229. CHM-Net: Center Heatmap-driven Macro-Micro Modeling Network for MRI-based Microbial Density Stratification ​

Author: Jiaming Liang, Haolin Chen, Tingting Li, Bowen Yu, Qianyan Long, Tinghe Zhang, Xi Zhong, Xiaowei Hu, Xiaoqi Sheng, Hongmin Cai
Published: 8/20/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2607.09812v2 Announce Type: replace-cross Abstract: Microbial density is clinically important for tumor assessment and treatment decision-making, and recent advances in deep learning suggest that it can be non-invasively inferred from multimodal MRI. In this work, MRI-based Microbial Density S...

📖 Read original article


230. From Adoption to Deployment: A Qualitative Study on AI Integration in Software Development Practice ​

Author: Mahzabin Tamanna, Elizabeth Lin, Sparsha Gowda, Laurie Williams, Dominik Wermke
Published: 8/20/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CR, cs.IR, cs.LG

arXiv:2607.16660v2 Announce Type: replace-cross Abstract: The increasing adoption of Large Language Models (LLMs) as AI components in modern software systems introduces distinct security risks to the software supply chain. While many considerations and safety mechanisms are in place for components o...

📖 Read original article


231. Constitutional Midtraining: Content Presence Drives Alignment Gains ​

Author: Desiree Cho, Cameron Tice, Bernie Hogan, Hunar Batra, Puria Radmard, Jun Zhao, Nigel Shadbolt
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG

arXiv:2607.26654v3 Announce Type: replace-cross Abstract: Post-training alignment is often shallow, eroding under fine-tuning. It remains untested as to whether constitutional midtraining interventions can produce durable alignment when cleanly isolated from post-training. We build a 394M-token cons...

📖 Read original article


232. Non-KKT Accumulation in Entropic Mirror Descent ​

Author: Kuangyu Ding, Kim-Chuan Toh
Published: 8/20/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DS

arXiv:2608.01658v3 Announce Type: replace-cross Abstract: For mirror descent generated by a Legendre kernel, perhaps one of the most basic question in optimization is this: must every accumulation point of a bounded mirror descent sequence be Karush--Kuhn--Tucker (KKT) stationary under proper stepsi...

📖 Read original article


233. Large-scale AI-Ready Data for Anti-Cancer Drug Response Modeling ​

Author: Vincent Lavelle, Yitan Zhu, Kaitlyn Marlor, Thomas Brettin, Rick Stevens
Published: 8/20/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG

arXiv:2608.11444v2 Announce Type: replace-cross Abstract: Drug response prediction (DRP) models are an active area of research in pharmacogenomics, with growing potential to accelerate the identification of effective anticancer drugs. However, their predictive performance is often constrained by lim...

📖 Read original article


234. Efficient Hessian-Free Methods for Multi-Objective Bilevel Optimization with Nonconvex Lower Level ​

Author: Yicong Jiang, Feihu Huang
Published: 8/20/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2608.12704v2 Announce Type: replace-cross Abstract: Multi-objective bilevel optimization has wide applications in the AI area such as automated learning and multi-task meta-learning. Although recently some works have been begun to study the multi-objective bilevel optimization, the proposed me...

📖 Read original article


235. Belayer: Efficient Fault Tolerance for LLM Agentic RL Training ​

Author: Jiecheng Zhou, Qinghao Hu, Peng Sun, Xingcheng Zhang, Weiming Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.14635v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents are increasingly trained with reinforcement learning in long-horizon, sandboxed environments. Unlike conventional RL, agentic RL couples GPU-intensive rollout engines with stateful environment containers whos...

📖 Read original article


236. RecurrentGPT: Expressive Depth through Recurrent Modulation in Transformers ​

Author: Amr Hegazy, Amr Alanwar, Mostafa Elhoushi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.15062v2 Announce Type: replace-cross Abstract: Scaling transformer language models creates an inherent tension between expressivity and memory efficiency. While unique weights across layers preserve functional specialization---from input-grounding to abstract refinement---they incur a sub...

📖 Read original article


237. The Null Token Knows: Reducing Message-Free Hallucination in ASR and NMT ​

Author: Kirill Borodin, Vasiliy Kudryavtsev, Ivan Viakhirev, Grach Mkrtchian
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SD

arXiv:2608.15940v2 Announce Type: replace-cross Abstract: Modern encoder-decoder systems can produce fluent text even when their input contains no recoverable message. We study this failure in ASR and NMT through the models' reserved null tokens, asking whether the score for ending generation alread...

📖 Read original article