arXiv cs.LG - 2026-08-21 ​
200 items collected.
1. Towards On-Board Implementation of ML-Based Helicopter Weight Estimator ​
Author: Nicolas Valot, Ammar Mechouche, Benjamin Lesage, Claire Pagetti, Louis Fabre
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.19210v1 Announce Type: new Abstract: This paper focuses on the implementation of a novel supervised Machine Learning model for estimating helicopter weight during takeoff, utilizing extensive datasets from Airbus's global in-service fleet. The study details a learning assurance process al...
2. Triangular Fuzzy Rescaling Distance ​
Author: Eddy Soria, Aida Valls, Ana Beatriz Hern'andez-Lara
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19234v1 Announce Type: new Abstract: Decision-making in complex systems often involves dealing with imprecise or uncertain information, frequently represented using fuzzy sets, particularly Triangular Fuzzy Numbers (TFNs). A crucial aspect of many fuzzy methods is the quantification of di...
3. Holtercare-Bench: A Multimodal Benchmark for Evaluating Long-Term Dynamic ECG Analysis ​
Author: Yihan Xie, Hanwen Cui, Runze Ye, Juekai Lin, Haoyang Wang, Jinhao Mao, Bo Zhang, Wenqiao Zhang, Xiaogang Guo, Jun Xiao, Lei Zhang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19297v1 Announce Type: new Abstract: While multimodal large language models (MLLMs) excel in medical applications, most of them favor static images or short-term signals. In the critical field of dynamic electrocardiograms (ECG), models struggle with complex temporal reasoning and diagnos...
4. Quantum Kernel Estimation for the Discovery of Early Lung Cancer Detection ​
Author: Hamed Javidi, Alex Zajichek, Hakan Doga, Laxmi Parida, Filippo Utro, Peter J. Mazzone
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.QM
arXiv:2608.19304v1 Announce Type: new Abstract: Lung cancer screening with low-dose chest computed tomography reduces mortality, but its impact is limited by uptake, adherence, and management challenges. Blood-based cell-free DNA (cfDNA) biomarkers offer a complementary approach, although early dete...
5. Improved Confidence Estimates for Black-Box Large Language Models ​
Author: Sokhna Diarra Mbacke, Mouloud Belbahri, Gabriel Loaiza-Ganem
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.19323v1 Announce Type: new Abstract: Uncertainty quantification (UQ) is essential for the safe deployment of large language models (LLMs). Existing methods, from verbalized confidence to ones requiring multiple generations, are often zero-shot and produce scores quantifying uncertainty wi...
6. Mechanistic Tomography: Designed Measurement for Control-Oriented Interpretability ​
Author: Vijay Erramilli
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.19338v1 Announce Type: new Abstract: Mechanistic interpretability seeks quantities that models do not expose directly: represented states, component effects, interactions, and responses to interventions. Patching, gradients, Hessian-vector products, and subset interventions provide differ...
7. Uncovering the Limits of Proof Sharing for Neural Networks ​
Author: Kanak Das, Shubham Ugare, Bor-Yuh Evan Chang, Sasa Misailovic, Gagandeep Singh, Manu Sridharan
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19351v1 Announce Type: new Abstract: Robustness verification of neural networks is increasingly important, due to their use in many critical domains. In certain scenarios, proof sharing has been shown to accelerate incomplete verification techniques by reusing intermediate-layer abstract ...
8. Longitudinal Bayesian Learning of Continuous Disease Position across the Alzheimer's Disease Continuum ​
Author: Yingying Zhang, Kun Zhao, Guodong Liu, Qi Huang, Pengfei Gu, Dongchul Kim, Erik Enriquez, Alex D. Leow, Paul M. Thompson, Heng Huang, Hongchang Gao, Liang Zhan, Haoteng Tang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.QM
arXiv:2608.19436v1 Announce Type: new Abstract: Alzheimer's disease (AD) progresses as a continuous biological process, whereas most existing neuroimaging-based artificial intelligence methods remain limited to discrete diagnosis or clinical score prediction from cross-sectional imaging. In this wor...
9. Quantifying Event Impacts on Time Series via Multiscale Contrastive Learning ​
Author: Yiming Sun, Shengyu Chen, Zhengzhang Chen, Haoyu Wang, Xiaowei Jia, Haifeng Chen
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19447v1 Announce Type: new Abstract: Shocks that spread through the web, such as cybersecurity breach disclosures, can abruptly disrupt financial time series and cause substantial abnormal losses. While these events are disclosed as discrete records through news reports, regulatory filing...
10. LLM as Detector: An In-context Learning Approach for Tabular Anomaly Detection ​
Author: Tu Anh Hoang Nguyen, Dang Nguyen, Thuc Duy Le, Trung Le, Sunil Gupta
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19463v1 Announce Type: new Abstract: Anomaly detection in tabular data is challenging because abnormal samples often arise as violations of cross-feature dependencies rather than simple marginal deviations. Existing detectors rely on geometric or reconstruction signals, while prior LLM-ba...
11. When to Retrain: An Empirical Study of Retraining Policies for Streaming ML Under Concept Drift, Budget, and Latency Constraints ​
Author: Sawan Dasari
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19488v1 Announce Type: new Abstract: Production machine learning systems degrade under concept drift, yet practitioners have little principled guidance on when to retrain. Retraining is costly, retraining budgets are finite, and a retrained model does not take effect instantly: training a...
12. DeltaMomentum: A Key-Value based Anisotropic Momentum Update via Delta Rule ​
Author: Euijin Hong, Guannan Qu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, math.OC, stat.ML
arXiv:2608.19491v1 Announce Type: new Abstract: Most modern optimizers form their momentum as an exponential moving average (EMA) of past gradients, forgetting every direction at one fixed rate. However, the inputs a deep network sees during training can be highly anisotropic, with a few directions ...
13. Beyond Multimodal Alignment: Certifying Physical Language through Response Substitution and Ordered Execution ​
Author: Kaizhen Tan, Xin Xu, Siru Tao, Yixiao Li, Hanzhe Hong, Yang Feng, Heqing Du
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2608.19492v1 Announce Type: new Abstract: World models increasingly treat compact multimodal representations as interfaces between perception and physical interaction, yet existing probes do not establish whether different sensors carry the same executable meaning or whether that meaning survi...
14. Empirical Characterization of Learning Geometry in Hybrid Quantum Forecasting Models ​
Author: Sandra Leticia Ju'arez-Osorio, Jorge I. Hernandez-Martinez, Jesus Ivan Ruiz-Martinez, Andres Mendez-Vazquez, Eduardo Rodriguez-Tello
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19497v1 Announce Type: new Abstract: We characterize the learning dynamics of a compact hybrid quantum forecasting model through comparison with a structurally aligned classical baseline. Using stationary harmonic-mixture and nonstationary chirp benchmarks with controlled spectral complex...
15. In Two Minds about Lifelong Learning: Exploring Hemispheric Redundancy and Specialisation in Neural Models ​
Author: Benjamin Smith, Levin Kuhlmann, Kaushik Roy, Gideon Kowadlo
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.19514v1 Announce Type: new Abstract: Persistent intelligent systems require the ability to learn continually, but current machine learning approaches face significant challenges in this area compared to biological learning systems. Machine learning algorithms typically trade off retention...
16. Continuous Adversarial MeanFlow Transfer ​
Author: Yara Bahram, Zahra Dehghani, M'elodie Desbos, Eric Granger, Pablo Piantanida, Mohammadhadi Shateri
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.19540v1 Announce Type: new Abstract: Training fast generators on new domains with limited data remains challenging for two reasons. First, adapting a pretrained diffusion or flow model to a new domain leaves its costly multi-step sampling unaddressed, and existing acceleration methods are...
17. DraftFM: A FoundationModel for Day-Zero Drafting in Magic: The Gathering ​
Author: Brian Ward
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.19568v1 Announce Type: new Abstract: Drafting a new Magic: The Gathering expansion begins before any pick from it has been observed: the complete card list is public, but the draft logs that supervised pick models train on do not yet exist. We study this day-zero regime directly. DraftFM ...
18. A Two-Stage Time-Aware Transformer for Short-Horizon AECOPD Risk Prediction ​
Author: Dongyang Wang, Weihao Qu, Ling Zheng, Haowen Pan
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19578v1 Announce Type: new Abstract: Acute exacerbation of chronic obstructive pulmonary disease (AECOPD) can worsen rapidly, making timely prediction a clinical priority. Most existing machine learning approaches rely on episodically collected clinical variables, introducing delays that ...
19. K\"ahler landscapes for complex neural network descents and guarantees including a search and destroy of the Calabi-Yau manifold ​
Author: Andrew Gracyk
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, math.DG, stat.ML
arXiv:2608.19584v1 Announce Type: new Abstract: We study landscapes for complex-parameterized networks. Our approach is motivated with an information-theoretic manifold perspective of the parameter and via classical optimization guarantees although of complex geometric variety such as through Dolbea...
20. Unregularized Convergence of Single-Loop, Entropy-Regularized Natural Actor-Critic ​
Author: Zhiqiang Tan
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19587v1 Announce Type: new Abstract: While entropy regularization is widely used to stabilize and accelerate Natural Policy Gradient methods, its ability to yield faster convergence rates for the unregularized objective remains underexplored. Existing analyses often rely on double-loop ar...
21. Complementary, Not Cumulative: Interaction Effects in Physics-Informed Neural Networks for Navier-Stokes Vortex Shedding ​
Author: Devesh Shah
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn
arXiv:2608.19632v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) embed governing partial differential equations directly into the training loss, offering a promising alternative to costly CFD solvers for unsteady flows. Yet the growing list of techniques proposed to improve P...
22. Time-Uniform Self-Normalized Concentration for Discounted Least Squares: Limits and Corrections ​
Author: Yi-Shan Wu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.19643v1 Announce Type: new Abstract: Self-normalized concentration inequalities are standard tools in bandit and reinforcement-learning analyses. A widely used weighted extension claims an analogous time-uniform guarantee for discounted least-squares estimators in non-stationary problems....
23. DeltaML-Bench: Evaluating Machine Learning Agents on Real-World Research Repositories ​
Author: Josias Moukpe, Priyanka Aryal, Matthew Kenney
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.19653v1 Announce Type: new Abstract: Autonomous agents for machine learning experimentation must navigate heterogeneous repositories, repair training pipelines, and evaluate candidate improvements under realistic compute constraints. Existing benchmarks only partially capture these condit...
24. Rationally Enriched Chebyshev Trunk Bases for DeepONet Surrogates of High P\'eclet Entrance Transport ​
Author: Mingeun Choi, Satish Kumar
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.19658v1 Announce Type: new Abstract: This study demonstrates a rationally enriched Chebyshev (REC) trunk for deep operator network (DeepONet) surrogate models of singularly perturbed and high-P'eclet transport problems whose solution profiles are characterized by thin localized boundary ...
25. FleetSieve: Decision-Critical Profiling for SLO-Aware LLM Fleet Configuration ​
Author: Huang Cheng, Scott Zhang, Aubert Li
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2608.19659v1 Announce Type: new Abstract: Choosing tensor-parallel (TP) degrees and replica counts for an LLM serving fleet is difficult because performance is not monotonic in TP and the feasible choice can change with load. Exhaustive profiling resolves this uncertainty, but measures many co...
26. SAGE-XGBoost: Spatially Augmented Graph Embeddings--Machine Learning Framework for Natural Hazards Susceptibility Mapping under Data Scarcity ​
Author: Mohammad H. Vahidnia, Ali Pourkarimi
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19672v1 Announce Type: new Abstract: Natural hazard susceptibility mapping is often constrained by limited labeled data, reducing the generalizability of conventional machine learning and limiting the applicability of complex deep learning models. This study proposes SAGE (Spatially Augme...
27. A Locally Tokenized Generative Model for Robust Time-Series Watermarking ​
Author: Dongbin Kim, Geonwoo Shin, Yujin Choi, Soyeon Park, Jaewook Lee
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.19727v1 Announce Type: new Abstract: Watermarking is a central tool for provenance in generative models, yet its application to multivariate time series remains hindered by reliability failures under post-editing attacks. We show that existing detectors, which rely on globally coupled re-...
28. RecPFN: Prior-Fitted Networks for In-Context-Based Recommendations ​
Author: En Zhi Tan, Jia Xiang Lim, Bryan Lijie Chew, Tze Minh Ng, Benjamin Yan Han Yap
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19735v1 Announce Type: new Abstract: We introduce RecPFN, a prior-fitted network that brings in-context learning to sequential recommendation. RecPFN is pretrained entirely on synthetic clickstream environments sampled from a broad structural causal prior, enabling it to amortize Bayesian...
29. Truncate Bad, Upweight Good: BoN-Style Distillation via Rank-Based Classification ​
Author: Yarin Bar, Yaniv Romano
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.19748v1 Announce Type: new Abstract: Inference-time selection methods, such as Best-of-N, improve generation by sampling a pool of candidates and selecting the top-ranked completion according to a reward model. Distillation seeks to amortize this procedure into a single policy by replacin...
30. Credit Without Ground Truth: Auditing Step-Level Credit Assignment in LLM Agents Against Executed Replay ​
Author: Haiyue Zhang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.19760v1 Announce Type: new Abstract: Audited against causal ground truth from executed replay in a single-agent tool environment (ALFWorld), none of the step-level credit signals used to train LLM agents -- LLM-judge scores, outcome-conditioned logprob ratios, or the policy's own confiden...
31. Finite-Horizon Input-Output Dynamics of Minibatch Perturbations in AdamW ​
Author: Kang Liu, Suyan Li
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC, stat.ML
arXiv:2608.19762v1 Announce Type: new Abstract: A minibatch can influence training beyond the update at which it is observed because AdamW stores past gradient information in its optimizer states. We study this delayed effect through paired trajectories that differ only in one gradient update and sh...
32. Unsupervised Anomaly Detection Using Flow Matching on Tabular Data ​
Author: Philip Konz, Tejaswini Medi, Margret Keuper
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19801v1 Announce Type: new Abstract: Financial anomaly detection often relies on large unlabeled transaction logs, where anomalous samples may already be present during training. Such training-set contamination violates the clean-normal data assumption underlying many anomaly detection me...
33. MileGPO: Milestone Inference with Local Evidence for Graph-Based Policy Optimization of Long-Horizon LLM Agents ​
Author: Bo Qian, Yuting Wu, Shuang Zeng, Huaiyu Wan, Dalin Zhang, Jiqiang Liu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.19803v1 Announce Type: new Abstract: Credit assignment is challenging in long-horizon agentic reinforcement learning, where supervision often comes only from final rewards. Existing methods refine trajectory-level signals into step-level credits through step grouping or graph-based advant...
34. Answer-Level Trust Selection for Physical Vision-Language Reasoning ​
Author: Rongyu Yu, Ke Niu, Fengxiang He
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19807v1 Announce Type: new Abstract: Vision-language models (VLMs) can estimate physical quantities such as duration, speed, and acceleration from visual observations, but existing benchmarks primarily assess overall model performance against annotated ground truth. In deployment, a key q...
35. FAR-DPO: Feasibility-Aware and Robust Direct Preference Optimization for Cyclic Peptide Design ​
Author: Guofeng Zhang, Rong Han, Xiaoyu Wang, Zhiyun Li, Zongbo Han, Xiaohong Liu, Guangyu Wang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19808v1 Announce Type: new Abstract: Cyclic peptides are emerging as promising molecular scaffolds in drug discovery due to their high binding affinity and structural stability. However, extending generative models from linear to cyclic peptide design remains challenging, as cyclization s...
36. Adaptive Probabilistic Shielding by Learning MDPs for Safe Reinforcement Learning ​
Author: Astrid Horn Brorholt (Aalborg University, Aalborg, Denmark), Maris F. L. Galesloot (Radboud University, Nijmegen, Netherlands), Nils Jansen (Radboud University, Nijmegen, Netherlands), Kim Guldstrand Larsen (Aalborg University, Aalborg, Denmark), Christian Schilling (Aalborg University, Aalborg, Denmark)
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.LO
arXiv:2608.19836v1 Announce Type: new Abstract: Probabilistic shielding is a technique for safe reinforcement learning (RL). Typically, a static observer -- called the shield -- constrains the learning agent's actions to those for which acting safely remains feasible. Traditionally, the shield is co...
37. Inadvertent Context Leakage in Language Models ​
Author: Jaiden Fairoze, Neal Mangaokar, Kamalika Chaudhuri, Sanjam Garg, Saeed Mahloujifar
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2608.19857v1 Announce Type: new Abstract: For AI agents to be useful beyond simple chat, they must hold sensitive user context such as calendars, credentials, health records, and financial data. We study whether the mere presence of such secrets in a model's context window introduces hidden co...
38. Online Test-Time Adaptation for Generalizable Dynamic Graph Anomaly Detection ​
Author: Jialun Zheng, Hanchen Yang, Jiannong Cao, Yankai Chen, Yuanjing Feng, Philip S. Yu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19858v1 Announce Type: new Abstract: Generalizable dynamic graph anomaly detection (DGAD) enables pretrained detectors to identify anomalies in unseen target domains without costly retraining. However, existing methods often fail for two reasons. First, they mainly rely on domain-agnostic...
39. Separating Covariate Shift from Mechanism Change with Two Discriminators: CJSD, a Conditional Discrepancy with an Exact Covariate-Concept Decomposition ​
Author: Kentaro Oda
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.19885v1 Announce Type: new Abstract: Streaming systems that maintain a pool of expert models must repeatedly decide whether to reuse an existing expert for arriving data, spawn a new one, or defer. We present a decision layer that makes all three outcomes statistically meaningful. Reuse a...
40. Evidence Before Expansion: Reuse, Spawn, or Defer in Lifelong Expert Pools ​
Author: Kentaro Oda
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2608.19888v1 Announce Type: new Abstract: Streaming systems that maintain a pool of expert models must repeatedly decide whether to reuse an existing expert for arriving data, spawn a new one, or defer. We present a decision layer that makes all three outcomes statistically meaningful. Reuse a...
41. Reliable Neural Collapse Approximation for Open-World Test-Time Adaptation ​
Author: Jia-Qi Lin, Yuangang Pan, Chang-Dong Wang, Haizhang Zhang, Ivor W. Tsang, Joey Tianyi Zhou
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19890v1 Announce Type: new Abstract: Test-Time Adaptation (TTA) methods aim to bridge the domain gap between the source and target domains. However, traditional TTA methods become ineffective when the label distribution shift occurs, a challenge commonly referred to as an open-world scena...
42. PETA:Parameter-Efficient Test-Time Adaptation for Virtual Screening ​
Author: Jia-Qi Lin, Yinghua Yao, Chang-Dong Wang, Yew-Soon Ong, Yuangang Pan
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19906v1 Announce Type: new Abstract: Accurately ranking active ligands for a target protein pocket from massive chemical libraries remains a central challenge in virtual screening. DrugCLIP and its recent extensions substantially accelerate this process by encoding protein pockets and mol...
43. Multi-Source Wasserstein Distributionally Robust Graph Learning ​
Author: Chuansen Peng, Yifan Xia, Jinshan Zhong, Xiaojing Shen
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19914v1 Announce Type: new Abstract: Network topology inference from graph signals is central to graph signal processing with applications in neuroscience, sensor, and social networks. In practice, target-domain samples are scarce while heterogeneous source-domain data are abundant. Fusin...
44. Auditing Recorded Predictive Lead Service-Line Classifications Against Physical Verification: A Statewide Study of New York ​
Author: Muhammad Sarmad Sohail
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CY
arXiv:2608.19922v1 Announce Type: new Abstract: Under the US Lead and Copper Rule Revisions, a utility may determine a service line's material with a predictive model instead of inspecting it. New York State publishes, per address, which method was used. Almost no address carries both a model classi...
45. G-MARK: Grounded Multi-Agent Reasoning for Cooperative Driving via Knowledge Graphs ​
Author: Bhavya Gupta, Onat Gungor, Tajana Rosing
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19964v1 Announce Type: new Abstract: Autonomous driving systems must operate under partial observability, where safety-critical objects may be occluded or visible only to neighboring connected vehicles. Vehicle-to-vehicle cooperation can reduce this uncertainty, but existing cooperative d...
46. Green BOA: Determining the environmental break-even point for ML-based data compression ​
Author: Caterina Doglioni, Akshat Gupta, Thomas Elliott, Hanzila Hussain, Sanjiban Sengupta
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, hep-ex, physics.comp-ph
arXiv:2608.19994v1 Announce Type: new Abstract: We summarise the outcome of two summer internship projects based at the University of Manchester, focused on the break-even point in terms of environmental sustainability for ML-based data compression algorithms. Using the example of a ML-based lossles...
47. Scale-Aware Pretraining of Time Series Foundation Models via Multi-Patch Token Alignment and Hybrid Masking ​
Author: Taihua Chen, Xiang Ma, Yixin Zhang, Tailin Zhan, Manyu Sun, Lizhen Cui
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20005v1 Announce Type: new Abstract: Pretraining time series foundation models across heterogeneous datasets necessitates effective handling of varying sampling frequencies. Current methods either employ dataset-specific patch sizes and separate FFNs, leading to fragmented representations...
48. Systematic Evaluation of TabPFN-TS for Zero-Shot Probabilistic Heat Load Forecasting in District Heating Networks ​
Author: Ben Spoek, Karim K. Ben Hicham, Kai Derzsi, Philipp Althaus, Alexander Mitsos, Dirk M"uller
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20024v1 Announce Type: new Abstract: District heating energy hubs require reliable heat load forecasts for efficient operational scheduling. Conventional forecasting workflows train system-specific models on historical data, which can become burdensome when networks change through new con...
49. CLaST: Context-aware Contrastive VAE for Probabilistic Time Series Forecasting ​
Author: Alexander Marusov, Dmitry Anikin, Petr Sokerin, Vitaliy Pozdnyakov, Ilya Kuleshov, Alexey Zaytsev
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20025v1 Announce Type: new Abstract: Probabilistic forecasting models are widely used for time series forecasting in domains such as energy systems, finance, medicine, and transportation. In recent years, deep generative models have shown strong results on probabilistic forecasting, yet m...
50. An Inclusive and Lightweight Approach to Federated Continual Learning for Cultural Heritage ​
Author: Ioannis Theologitis, Debin Meng, Stylianos Eleftheriadis, Vasileios Lolis, Konstantinos Votis
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.20038v1 Announce Type: new Abstract: Artificial intelligence can support cultural heritage and digital humanities through large-scale retrieval and analysis of digitized collections. However, cultural heritage data are often distributed across institutions, constrained by ownership and ac...
51. End-to-end Early Classification of Time Series in Non-Stationary Environments ​
Author: Aur'elien Renault, Alexis Bondu, Antoine Cornu'ejols, Vincent Lemaire
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20044v1 Announce Type: new Abstract: Early Classification of Time Series (ECTS) requires making accurate decisions as early as possible in inherently online and evolving environments. Yet, most existing methods assume stationarity and rely on separable designs, where classification and tr...
52. DecoVAE: a Lightweight Interpretable Trend-Seasonal VAE Framework for Efficient Probabilistic Time Series Forecasting ​
Author: Alexander Marusov, Dmitry Anikin, Alexey Zaytsev
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20052v1 Announce Type: new Abstract: Probabilistic time series forecasting remains challenging, largely because modeling distinct trend and seasonal dynamics requires specialized approaches. Existing methods often fail to capture the unique inner properties of these components, lack inter...
53. Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts ​
Author: Nayeon Kim, Hojin Lee, Yunju Bak, Jaesun Park, Boseop Kim
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.20061v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures significantly expand model capacity without a proportional increase in computational cost. However, optimizing their hyperparameters---particularly the learning rate---at extreme scales of both model size and toke...
54. Orthogonal JEPA: Factorized Predictive States for Latent World Models ​
Author: Taoyong Cui, Pheng Ann Heng, Wanli Ouyang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20065v1 Announce Type: new Abstract: World models construct latent states that support prediction, planning, and reasoning about an underlying system. Joint-embedding predictive architectures (JEPAs) offer a direct way to learn such states by predicting targets in representation space ins...
55. SAE-Xplainers: Rule-Based Feature Interpretation for Extreme Earth Events ​
Author: Hugo Porta, Emanuele Dalsasso, Chang Xu, Theo Gnassounou, Devis Tuia
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20117v1 Announce Type: new Abstract: The emergence of large-scale Weather and Climate (W&C) datasets offers new opportunities for modeling extreme Earth events (ExEE) and their impacts using deep learning. However, their adoption in operational settings remains limited by the lack of mode...
56. Evaluating Neural Cartographic Relief Shading for Urban Environments: A Downtown Calgary Study Using High-Resolution DEM and DSM Data ​
Author: Emmanuel Stefanakis
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20149v1 Announce Type: new Abstract: This article explores the performance of analytical and neural-based hillshading methods in a dense urban environment using high-resolution digital elevation model (DEM) and digital surface model (DSM) data for downtown Calgary. The study compares sing...
57. Ask Self, Ask Others: Relation Is All You Need ​
Author: Yuting Ge, Pengju Yang, Mingkai Nie
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20172v1 Announce Type: new Abstract: Attention directly derives normalized information flow from pairwise scores. We introduce Relation, an alternative token-mixing primitive that first organizes pairwise evidence into explicit Self and Exchange relations and derives information flow afte...
58. A Standardized Framework for Machine Learning in Power System Protection ​
Author: Julian Oelhaf, Georg Kordowich, Paula Andrea P'erez-Toro, Christian Bergler, Johann J"ager, Andreas Maier, Siming Bayer
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP
arXiv:2608.20181v1 Announce Type: new Abstract: Studies of machine-learning-based power-system protection increasingly report near-perfect scores, yet the meaning of those scores depends strongly on the evaluation setting. Protection task, physical scope, measurements, timing, targets, preprocessing...
59. Exact Algebraic Computation of Learning Coefficients for Two-Dimensional Singular Models ​
Author: Gr'egoire Sergeant-Perthuis (CQSB, Sorbonne Universit'e), Elias Tsigaridas (Ouragan Team, INRIA), Jules Tsukahara (Ouragan Team, INRIA)
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.SC, math.AG, stat.ML
arXiv:2608.20183v1 Announce Type: new Abstract: Classical information criteria such as the Bayesian Information Criterion (BIC) rely on regularity assumptions that break down for singular models, leading to incorrect model selection in settings such as deep learning. The Widely Applicable Bayesian I...
60. Decoding silent reading from non-invasive EEG ​
Author: Ingo Marquardt, Anthilia Alchanat, Priyanka Jain
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC
arXiv:2608.20186v1 Announce Type: new Abstract: Non-invasive decoding of inner speech faces a fundamental data problem: a corpus pairing brain activity with a person's spontaneous inner monologue cannot be collected, and the available proxy paradigms (cued repetitive and retrospectively reported gen...
61. DICS: Data-Informed Centroid Splitting for Decision Tree Classifiers ​
Author: MD Saifur Rahman Mazumder, Feng Yu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.20258v1 Announce Type: new Abstract: Decision tree-based models are widely used in machine learning due to their interpretability and strong empirical performance. However, training decision trees can be computationally expensive, particularly for large and high-dimensional datasets, larg...
62. Dynamic Structural Causal Modeling for Sleep ​
Author: Ranveer Singh, Saurabh Mathur, Pranuthi Tenali, Arun Badi, Sriraam Natarajan
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20285v1 Announce Type: new Abstract: The causal dynamics of sleep-disordered breathing are complex and vary across patient populations, hindering the development of targeted interventions. We learn dynamic causal graphs of sleep-disordered breathing from Home Sleep Apnea Test (HSAT) recor...
63. Physical-Support Confidence Sets for Highly Coherent Dictionaries ​
Author: Guan-Ju Peng
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, math.ST, stat.TH
arXiv:2608.20295v1 Announce Type: new Abstract: Sparse pursuit after dictionary learning can yield a precise atom support even when its physical interpretation is not justified by the calibration data, especially for highly coherent dictionaries where alternative calibration-compatible dictionaries ...
64. Explainable Transformer Models for Clinical Prediction Tasks on Structured Electronic Health Records ​
Author: Jun Ni Du, Lukas Adamek, Maxim Kryukov, Flavio Dormont, Ziv Bar-Joseph, Sven Jager, Brandon Rufino
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20315v1 Announce Type: new Abstract: Predictive models over structured electronic health records (EHRs) remain central to machine learning for healthcare, but few have jointly emphasized quantitative laboratory information and interpretability with respect to input medical events. We pres...
65. A comparison between ceiling-mounted FMCW, IR-UWB and Wi-Fi radar for in-bedroom human activity monitoring and sleep interruption detection ​
Author: Anton Lambrecht, Reda El Hail, Xianjun Jiao, Pieter Crombez, Dominique Schreurs, Peter Karsmakers, Adnan Shahid, Eli De Poorter
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20322v1 Announce Type: new Abstract: Despite their growing importance for contact-free radio frequency (RF) based healthcare monitoring, different radio technologies such as frequency-modulated continuous wave (FMCW) radar, impulse radio ultra-wideband (IR-UWB), and Wi-Fi sensing are rare...
66. VQC-ZTI: Variational Quantum Control for Zero Trust Protection of the Tactile Internet ​
Author: Mubassir Serneabat Sudipto (Iowa State University), Shakil Ahmed (Grand Valley State University), Ashfaq Khokhar (Kansas State University)
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.NI
arXiv:2608.18572v1 Announce Type: cross Abstract: Tactile Internet services couple cyber events directly to physical actuation, so security decisions must improve risk discrimination without perturbing the control path. This paper presents VQC-ZTI, a split-plane Variational Quantum Classifier framew...
67. Active Inference as Context Acquisition for AI Agents ​
Author: Sanchayan Dutta, Sai Niranjan Ramachandran, Suvrit Sra
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.19202v1 Announce Type: cross Abstract: Interactive AI agents must acquire the right context as efficiently as possible. When a user omits a constraint, preference, file, or task variable, an agent can proceed with a default assumption or spend tokens on a clarifying question, retrieval ca...
68. Asymmetric Attention Heads: Structured Head-Wise Context Allocation for Transformer Attention ​
Author: Zimu Zhao
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.19203v1 Announce Type: cross Abstract: Standard multi-head attention (MHA) gives every head the same full causal context span, although heads can serve different contextual roles. Some heads may rely mainly on nearby lexical or syntactic context, while others may depend on longer-range re...
69. Time-Series Retrieval for Grounding Multimodal Language Models in Remaining Useful Life ​
Author: Valeriu Dimidov, Rapha"el Frank
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.19218v1 Announce Type: cross Abstract: Large language models (LLMs) and agentic AI systems are increasingly being explored for domain-specific maintenance and prognostics tasks, raising the question of whether they can effectively support prognostics and health management (PHM). In this p...
70. Causal Inference under Interference with Learned Exposure Mappings ​
Author: Cong Cao
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ME, cs.AI, cs.CY, cs.LG, stat.AP
arXiv:2608.19224v1 Announce Type: cross Abstract: Exposure mappings are often assumed to be known in causal spillover analyses. In environmental settings, however, they are typically induced by transport processes that are not directly observed and must instead be learned from pollution data. We stu...
71. M3: A State-Event Generative Foundation Model for Market Microstructure Dynamics ​
Author: Yanzhi Zhang, Yu Ma, Yilin Cheng, Jian Li, Yitong Duan
Published: 8/21/2026, 4:00:00 AM
Categories: q-fin.CP, cs.LG
arXiv:2608.19227v1 Announce Type: cross Abstract: Market microstructure simulation aims to model how liquidity, prices, and order flow evolve in electronic financial markets. Since market data reveal only one realized trajectory, many important questions are inherently counterfactual and require rea...
72. TorchDCM: A Unified PyTorch-Native Package for Discrete Choice Modeling ​
Author: Baichuan Mo, Zhengzhong Ricky You, Xiqun Michael Chen, Ruimin Li
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.19231v1 Announce Type: cross Abstract: Estimating large and simulation-intensive discrete choice models (DCMs) requires repeated evaluation of utilities, probabilities, derivatives, and simulated likelihoods over many observations, alternatives, and draws. Existing DCM software provides m...
73. Active Spiking Perception: The Membrane Potential as a Belief State for Anytime 3D Point Cloud Recognition ​
Author: Akarsh Jain, Arya Pawa, Ayush Debnath, Smera Rawal, Sayeed Shafayet Chowdhury
Published: 8/21/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG
arXiv:2608.19232v1 Announce Type: cross Abstract: Spiking point cloud networks usually scan space in a fixed, input-agnostic order, which leaves the most distinctive resource of spiking computation, the temporal evolution of the membrane potential, unused as a locus of decision-making. Active Spikin...
74. Demons on a Budget: Adaptive Measurement Placement at the Entanglement Phase Transition ​
Author: Rohan Pandey
Published: 8/21/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.19248v1 Announce Type: cross Abstract: Monitored quantum circuits exhibit a measurement-induced phase transition between volume-law and area-law entanglement as a function of the measurement rate $p$. Prior work places measurements at random locations and treats the rate as the control pa...
75. Recovering Nonlinear Functions of Latent Variables: A Plausible-Value Neural Network Framework ​
Author: Eunjeong Song (Department of Education, Korea University, Seoul, Republic of Korea), Sehee Hong (Department of Education, Korea University, Seoul, Republic of Korea)
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ME, cs.LG
arXiv:2608.19282v1 Announce Type: cross Abstract: When factor scores replace true latent scores in nonlinear prediction, measurement error attenuates the recoverable variance of any $k$th-order component of the regression function by $\rho^k$ -- the $k$th power of the score's coefficient of determin...
76. Clustering and Token Denoising for Faster and More Robust VLMs ​
Author: Baptiste Rossigneux, Inna Kucher, Vincent Lorrain, Emmanuel Casseau
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.19285v1 Announce Type: cross Abstract: Recent Visual-Language Models (VLMs) have enhanced the capabilities of pre-trained LLMs by adding vision tokens alongside text, with approaches like LLaVA showing impressive results. However, the computational burden of processing up to 576 or 729 vi...
77. Quantum Gaussian processes for prediction of channel observations ​
Author: Jonas J"ager, Yaroslav Khmelnitskiy, Paolo Braccia, Artur Miroszewski, Diego Garc'ia-Mart'in, M. Cerezo, Piotr Czarnik
Published: 8/21/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, stat.ML
arXiv:2608.19306v1 Announce Type: cross Abstract: Given a set of input states, we consider the task of predicting the expectation value of a Pauli observable at the output of an unknown quantum evolution, using only a limited number of measurements. Recently, quantum Gaussian process (QGP) regressio...
78. Deep neural networks as lattice gauge theories ​
Author: Ro Jefferson, Shradha Ramakrishnan
Published: 8/21/2026, 4:00:00 AM
Categories: hep-th, cond-mat.dis-nn, cs.LG
arXiv:2608.19331v1 Announce Type: cross Abstract: We modify the NN/QFT duality [1] to incorporate the layerwise permutation symmetry of the network, resulting in a $(0!+!1)$-dimensional lattice gauge theory, in which each layer of $N$ neurons acts as an $N$-component lattice site, and the weight m...
79. Data-Driven Time-Varying Control Barrier Functions for Adaptive Safe-Set Learning with Online Decremental Support Vector Machines ​
Author: Shawon Dey, Michael Budihartono, Hever Moncayo
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CE, cs.LG
arXiv:2608.19366v1 Announce Type: cross Abstract: Mission-critical intelligent systems often operate under time-varying limitations that reduce control authority and change the admissible safe operating envelope. In such settings, a safety certificate learned under nominal conditions may become inva...
80. Heteroscedastic Neural Surrogate Modeling for Robust and Rapid Bayesian Inference in Fusion Plasma Diagnostics ​
Author: Liyun Zhang, Naoya Mamada, Kentaro Sakai, Takeo Hoshi, Toru Aonishi
Published: 8/21/2026, 4:00:00 AM
Categories: stat.CO, cs.LG
arXiv:2608.19377v1 Announce Type: cross Abstract: Bayesian inference via Markov Chain Monte Carlo (MCMC) provides effective parameter estimation, but its real-time application in complex physical systems is hindered by heavy computational bottlenecks and extreme sensitivity to statistical noise. We ...
81. Concentrated Liquidity Provision: a Reinforcement Learning Perspective ​
Author: Georgios Chionas, Charalampos Kleitsikas, Stefanos Leonardos, Leandro S'anchez-Betancourt, Carmine Ventre
Published: 8/21/2026, 4:00:00 AM
Categories: q-fin.TR, cs.AI, cs.LG, q-fin.CP, q-fin.MF
arXiv:2608.19389v1 Announce Type: cross Abstract: Automated market makers (AMMs) are a cornerstone of decentralised finance (DeFi). Constant product markets with concentrated liquidity, such as UniswapV3, are now a well-established design. In these markets, liquidity providers (LPs) face a sequentia...
82. Deep-MKV-TS: Path-Dependent McKean--Vlasov Control for Financial Time Series Generation ​
Author: Samer El Boustany, Th'eo Basseras, Samy Mekkaoui, Alexandre Alouadi, Yadh Hafsi, Huy^en Pham
Published: 8/21/2026, 4:00:00 AM
Categories: q-fin.CP, cs.CE, cs.LG, math.OC
arXiv:2608.19394v1 Announce Type: cross Abstract: We introduce Deep-MKV-TS, a path-dependent McKean-Vlasov framework for financial scenario generation. The stochastic dynamics are chosen by matching selected path and volatility features of generated scenarios to those observed in the data. Starting ...
83. HYDRA: A Heterogeneous Chiplet DSE Framework for Serving Dynamic Hybrid LLM Workloads ​
Author: Jiahao Lin, Alish Kanani, Sangwan Lee, Jaehyun Park, Umit Ogras
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AR, cs.AI, cs.LG
arXiv:2608.19395v1 Announce Type: cross Abstract: Hybrid Transformer-Mamba large language models (LLMs) enhance long-context efficiency, but their heterogeneous computation and communication patterns complicate efficient hardware acceleration. Chiplet-based architectures offer a scalable solution by...
84. HiRA-CAM: Preserving Fine-Grained Spatial Relevance in Gradient-Based Visual Explanations ​
Author: Manasi Nerurkar, Ali A. Minai
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.NE
arXiv:2608.19407v1 Announce Type: cross Abstract: Deep Learning models can include billions of parameters or more, making it difficult to explain their internal transformations and outputs. However, explainability is increasing in importance due to the use of AI in crucial applications. This paper f...
85. Beyond Imitation: Filtering On-Policy Distillation by Reasoning Progress ​
Author: Chen Yang, Haiyuan Wan, Rengrong Xiong, Yize Chen, Danny H. K. Tsang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.19408v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as an effective framework for post-training language models by pairing student-generated trajectories with dense token-level supervision from a teacher. However, OPD implicitly assumes that teacher-derived rew...
86. Microlensify: a Transformer Based Machine Learning Classifier for Microlensing Events Trained on TESS Light Curves ​
Author: Atousa Kalantari, Somayeh Khakpash, Sedighe Sajadian, Hosein Haghi, Willow Fox Fortino, Rosanne Di Stefano
Published: 8/21/2026, 4:00:00 AM
Categories: astro-ph.IM, astro-ph.EP, astro-ph.GA, astro-ph.SR, cs.LG
arXiv:2608.19419v1 Announce Type: cross Abstract: Microlensing can reveal populations of faint compact objects that are otherwise difficult to detect. Depending on their design, all-sky surveys have the potential to search for these objects across the sky. The Transiting Exoplanet Survey Satellite (...
87. SCAPE: Scenario-Conditioned Simulation-Augmented Policy Evaluation ​
Author: Dijie Zhu, Seunghun Oh, Ruopeng Huang, Zhiyu Huang, Jiaqi Ma, Chen Tang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2608.19425v1 Announce Type: cross Abstract: Reliable performance evaluation is a central bottleneck for deploying robot-learning policies in real-world conditions. Real-world testing is faithful but costly and difficult to scale, whereas simulation-based testing scales easily but is inevitably...
88. Fine-Tuning VLAs with Self-Demonstrated Generative Control for Multi-Task Manipulation ​
Author: Prachi Garg, Steve Xing, Prahit Yaugand, Saurabh Gupta, Derek Hoiem
Published: 8/21/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG
arXiv:2608.19490v1 Announce Type: cross Abstract: State-of-the-art vision-language-action (VLA) models such as $\pi_{0.5}$ exhibit strong semantic understanding, instruction following and task behavior. However, when deployed on new robots, even minor mismatches in hardware configuration relative to...
89. Composition-Driven Phase Evolution in Sm-Doped BiFeO3 via Latent-Field Reconstruction of Atomically Resolved STEM Data ​
Author: Newsha Javanmardi, Christopher T. Nelson, Anna N. Morozovska, Eugene A. Eliseev, Ichiro Takeuchi, Sergei V. Kalinin
Published: 8/21/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG
arXiv:2608.19544v1 Announce Type: cross Abstract: Functionalities of ferroelectric materials are governed by the spatial organization and coupling of polarization, strain, lattice rotation, and structural order accessible via atomically resolved scanning transmission electron microscopy (STEM) image...
90. Forking Fast: Efficiently Estimating Uncertainty Dynamics in Text Generation ​
Author: Eric Bigelow, Amir Zur, Satchel Grant, Tal Haklay, Can Rager, Owen Lewis, Thomas McGrath, Jack Merullo, Ekdeep Singh Lubana, Atticus Geiger
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.19611v1 Announce Type: cross Abstract: LLM reasoning is stochastic, and so understanding a model requires grappling with the distribution of reasoning chains that it might produce for a given question, i.e., its uncertainty. Resampling-based analyses characterize this distribution, reveal...
91. Scaffolding Minds: Optimizing Latent Visual Target Representations for Multimodal Reasoning ​
Author: Haoqiang Kang, Yinpeng Chen, Luyang Liu, Jesper Sparre Andersen, Abhijit Ogale, Baochen Sun, Lichan Hong, Ed H. Chi
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.19669v1 Announce Type: cross Abstract: Latent reasoning has advanced multimodal reasoning through a two-stage training paradigm: (1) a helper image is encoded into latent tokens to teach visual chain-of-thought during a supervised fine-tuning (SFT) stage, and (2) these latent tokens are f...
92. CacheRoute: Planned Prefix-Affinity Routing for Large-Scale LLM Serving ​
Author: Huang Cheng
Published: 8/21/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2608.19677v1 Announce Type: cross Abstract: Prefix caching avoids prefill only when a repeated request returns to a server that still holds the prefix KV. Cache-blind balancing disperses that reuse; fixed affinity preserves it but can overload a server. CacheRoute resolves this tradeoff with a...
93. Learning Hierarchical Skill Policies with Offline Quality-Diversity Reinforcement Learning ​
Author: Tanachai Anakewat, Takayuki Osa, Tatsuya Harada
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.RO
arXiv:2608.19684v1 Announce Type: cross Abstract: Recent studies investigate how to leverage pre-collected datasets to improve the policy performance and sample efficiency of RL. One promising approach to achieve this goal is to employ a two-stage strategy: In the first stage, diverse skills are ext...
94. Learning Deterministic and Stochastic Forced Hamiltonian Systems ​
Author: Benedikt Brantner, Tomasz Tyranowski
Published: 8/21/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.SG
arXiv:2608.19688v1 Announce Type: cross Abstract: We develop a geometric framework for learning deterministic and stochastic forced Hamiltonian systems with neural networks. Motivated by the Lagrange-d'Alembert principle and the theory of variational integrators, we introduce the notion of a Lagrang...
95. RIPE++: Reinforced Keypoint Learning from Positive Pairs Only ​
Author: Johannes K"unzel, Peter Eisert, Anna Hilsmann
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.19693v1 Announce Type: cross Abstract: Sparse keypoint extraction and matching underpin core tasks in geometric computer vision, including structure-from-motion, visual SLAM, augmented reality, and medical image registration. Learning robust local feature representations, however, typical...
96. Projector Is All You Train ​
Author: Nyx Iskandar, Saathvik Selvan, Slater Victoroff
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.CV, cs.LG
arXiv:2608.19726v1 Announce Type: cross Abstract: The typical training process of a multimodal large language model (MLLM) involves adapting both the language model backbone and the projector between the backbone and a modality-specific encoder. We ask whether fine-tuning the backbone of an MLLM is ...
97. Question-Guided Evidence Acquisition for Multimodal Visual Question Answering ​
Author: Alin-Ionut Popa
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.19739v1 Announce Type: cross Abstract: Multimodal LLMs can see a document, but they often can't read it reliably. Small text, tables, visual cues, and topological elements still trip them up under direct visual inference, even when the page is already sitting in the model's context. Most ...
98. Far from the Crowd: Scalable Self-Supervised Learning via Geographic Isolation ​
Author: Daniele Rege Cambrin, Francesco Rossi, Mattia Varile
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.19766v1 Announce Type: cross Abstract: Self-supervised pretraining on remote sensing imagery typically treats all samples as equally informative, despite large variability in geographic and visual structure. We propose a curriculum learning strategy for self-supervised Earth observation t...
99. skchange: Fast and Flexible Algorithms for Changepoint Detection ​
Author: Martin Tveten, Johannes Voll Kolst{\o}, Per August Jarval Moen
Published: 8/21/2026, 4:00:00 AM
Categories: stat.CO, cs.LG
arXiv:2608.19767v1 Announce Type: cross Abstract: Skchange is an open-source Python library for detecting structural changes in time series. It implements modern change detection algorithms within a unified and extensible framework. The algorithms are modular and composable, and they include changep...
100. An Irreducible Quantum Advantage in Aligning World Models with Reality ​
Author: Josep Lumbreras, Hailan Ma, Jayne Thompson, Mile Gu
Published: 8/21/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG
arXiv:2608.19779v1 Announce Type: cross Abstract: World models provide digital simulacra of the true world, allowing agents to be trained and tested before costly real-world deployment. At each time step, they receive an action and generate an observation and reward matching the statistics of the tr...
101. Learning piecewise-smooth dynamical systems ​
Author: Davide Murari, Erik Jansson, Chris Budd OBE, Carola-Bibiane Sch"onlieb
Published: 8/21/2026, 4:00:00 AM
Categories: math.DS, cs.LG, cs.NA, math.NA
arXiv:2608.19785v1 Announce Type: cross Abstract: Discovering dynamical systems from trajectory data is a central problem in applied mathematics and engineering. Whilst recent advances in machine learning have led to strong progress in data-driven system identification, much less attention has been ...
102. PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents ​
Author: Seongjae Kang, Taehyung Yu, Sung Ju Hwang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.19861v1 Announce Type: cross Abstract: Customer-service LLM agents must follow organizational policy when acting on a user's behalf. Compliance failures arise from either forbidden actions, such as granting an ineligible change, or omitted procedural requirements, such as identification o...
103. A Repeated Measurements Approach to $SoH$ Battery Modelling of Cyclic Aged Data in a Laboratory Environment ​
Author: Mark Cary, Charles Bokor
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, cs.SY, eess.SY
arXiv:2608.19879v1 Announce Type: cross Abstract: This document describes the application of a first order linearised nonlinear repeated measurements approach to the analysis of battery cell ageing profiles generated under controlled conditions in a laboratory. The primary advantage of the model is ...
104. EnvHarness: Awakening Static Worlds for Agent Learning ​
Author: Chengsong Huang, Zifeng Wang, Rujun Han, Jun Yan, Yanfei Chen, Zoey CuiZhu, Ke Jiang, Peng Xia, Han Yu, Yufan Zhuang, Yifei Ming, Jiaqi Pan, Bhavana Dalvi Mishra, Jiaxin Huang, Burak Gokturk, Tomas Pfister, Chen-Yu Lee
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.19880v1 Announce Type: cross Abstract: LLM agents learn by interacting with environments, yet these environments are hand-built and static: blind to an agent's weaknesses, and quickly left behind as it improves. While recent environment generation methods attempt to address this, they req...
105. Interpretable Feature Learning for RF Fingerprinting via Polar MKANs ​
Author: Mikhail Krasnov, Ljupcho Milosheski, Carolina Fortuna
Published: 8/21/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.19881v1 Announce Type: cross Abstract: Radio frequency (RF) fingerprinting authenticates wireless devices from hardware-induced I/Q impairments, typically with deep learning feature extractors that are accurate but opaque, limiting their use in security critical settings. We propose Polar...
106. The impact of feature engineering and an optimisation framework for ocean colour machine learning ​
Author: Edson Silva, Julien Brajard, Simon Cappe, Lasse H. Pettersson, Fran\c{c}ois Counillon
Published: 8/21/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG
arXiv:2608.19899v1 Announce Type: cross Abstract: Machine learning (ML) is widely used for the development of ocean colour algorithms, but most studies focus on model parameter training and hyperparameter tuning. The optimisation of the data that feeds the models - i.e., Feature Engineering (FE) - i...
107. Where Does the Union Bound Go? Best-Arm Identification and Strong FWER Control ​
Author: Rianne de Heide
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML
arXiv:2608.19903v1 Announce Type: cross Abstract: In fixed-confidence best-arm identification, proofs often use a union bound across the competing arms. From a multiple-testing point of view this can look puzzling: if the best arm is unique, only one hypothesis of the form ``arm $i$ is best'' can be...
108. Spike-based Belief Propagation in Nonlinear Dynamical Systems ​
Author: Sepideh Adamiat, Hongye Wang, Wouter M. Kouw, Bert de Vries
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NE, cs.SY, eess.SY
arXiv:2608.19907v1 Announce Type: cross Abstract: This paper presents a Bayesian control framework that integrates spike-based dynamics with probabilistic inference for adaptive control. Bayesian inference is widely regarded as a core computational principle of brain function, providing a normative ...
109. A Layered Simplex Architecture for Large Alphabets ​
Author: Meir Feder, Yaniv Fogel, Ruediger Urbanke
Published: 8/21/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT, stat.ML
arXiv:2608.19908v1 Announce Type: cross Abstract: Probability estimation over large alphabets under log loss is a well-studied problem, with celebrated methods such as the Good-Turing estimator. We introduce and study a new Bayesian estimator with four notable properties. First, its construction is ...
110. From Noise to Signal: Improving Security Log Anomaly Detection Using LLMs with Endpoint-Specific Logs ​
Author: Christopher Henshaw, Gour Karmakar
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.19938v1 Announce Type: cross Abstract: Existing approaches to anomalous behaviour log detection, such as Wazuh rely primarily on predefined detection rules, while statistical anomaly detection approaches such as OpenSearch identify deviations from previously observed behavioural patterns....
111. Flow Matching Meets 3D Curvilinear Structure Segmentation in Medical Imaging ​
Author: Sidi Mohamed Sid'El Moctar, Nicolas Vitry, H'el`ene Bouvrais
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, q-bio.QM
arXiv:2608.19965v1 Announce Type: cross Abstract: Segmentation of curvilinear anatomical structures in 3D medical images remains challenging due to complex topology, severe class imbalance, weak contrast, and large variations in structure morphology. While deep learning approaches for 3D curvilinear...
112. From Street View Imagery to Street Quality Indicators: Vision Language Inference for the Suburban 15-minute City ​
Author: Joan Perez, Giovanni Fusco
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.20026v1 Announce Type: cross Abstract: Streetscape quality has become a central concern in contemporary urban planning, particularly within the framework of the pedestrian-friendly 15-minute city, where walkability and public-space quality are increasingly recognized as key determinants o...
113. Auditing Cross-Lingual Fairness in Language Model Watermarking ​
Author: Alexander Nemecek, Osama Zafar, Debargha Ganguly, Vikash Singh, Vipin Chaudhary, Erman Ayday
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.CR, cs.LG
arXiv:2608.20047v1 Announce Type: cross Abstract: Watermarking schemes for large language model output are evaluated almost exclusively on English text using each scheme's detection threshold and a narrow set of quality measurements. Multilingual deployment exposes evaluation-design choices that are...
114. What You Can't See Is What You Learn: Restricted Evidence Visibility Favors Compositional Generalization in Shared-Genome Language-Model Societies ​
Author: Narcis Marincat
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA
arXiv:2608.20054v1 Announce Type: cross Abstract: Multi-module systems often expose every module to the full input. We test whether restricting evidence visibility changes which solutions gradient-based training discovers. Four-cell societies share one frozen pretrained language model and one low-ra...
115. Reward-Guided Autoregressive Graph Generation for Efficient Multi-Agent Communication Topology Design ​
Author: Poomphob Suwannapichat, Boonyarit Changaival, Caesar Wu, Pascal Bouvry
Published: 8/21/2026, 4:00:00 AM
Categories: cs.MA, cs.CL, cs.LG
arXiv:2608.20099v1 Announce Type: cross Abstract: LLM-based Multi-Agent Systems (MAS) achieve strong performance on complex reasoning tasks by coordinating multiple agents, but at the cost of substantial token consumption. Recent work on automatic topology design, ARG-Designer, has reframed this pro...
116. Structured Affinity for Unsupervised Visual Class-Incremental Memory in Deep Artificial Immune Networks ​
Author: Siphesihle Sithungu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.20104v1 Announce Type: cross Abstract: Artificial immune networks (AINs) are naturally memory-forming systems, but conventional visual AINs often rely on flattened vector affinity that ignores spatial structure. This paper studies whether structured, gradient-free immune affinity can make...
117. Discrete Diffusion Inference-Time Control with Nested Sequential Monte Carlo ​
Author: Lohithsai Yadala Chanchu, Hany Abdulsamad, Christian A. Naesseth
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.20123v1 Announce Type: cross Abstract: We study inference-time control for text generation in discrete diffusion language models, where the goal is to steer sampling toward sequence-level rewards without retraining. Prior work in this domain has focused on particle-based methods such as b...
118. Feature Evolution and Migration during Vision Transformer Training ​
Author: Joonas J"arve, Halil Ibrahim Aysel, Tarun Khajuria, Meelis Kull
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.20134v1 Announce Type: cross Abstract: We present a novel view on feature evolution in Vision Transformers (ViTs) by visualizing the training process over two dimensions -- network depth (layer) and training time (epochs). We employ Sparse Autoencoders (SAEs) to extract candidate sparse f...
119. Task-CoEvolve: Efficient Harness Optimization via Adaptive Validation Task Selection ​
Author: Atsuyuki Miyai, Kiyoharu Aizawa, Toshihiko Yamasaki
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.20169v1 Announce Type: cross Abstract: We present a novel approach to efficient LLM agent harness optimization through adaptive validation task selection. Harness optimization iteratively rewrites the harness code based on validation performance, enabling substantial performance gains wit...
120. MemTrapBench: Benchmarking Cognitive Traps in LLM Memory Use ​
Author: Mengru Wang, Haozhe Luo, Zhenqian Xu, Zhixiang Cui, Haoming Xu, Qu Yang, Jizhan Fang, Junfeng Fang, Ningyu Zhang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CY, cs.DB, cs.LG
arXiv:2608.20202v1 Announce Type: cross Abstract: Memory has become a key component of large language models, enabling them to retain information and learn from long-term interactions. However, existing memory benchmarks mainly evaluate whether information is correctly extracted, stored, and retriev...
121. Daedalus-150M: A Convolution-Attention Hybrid Designed for CPU Inference ​
Author: Christos Koutsiaris
Published: 8/21/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG
arXiv:2608.20210v1 Announce Type: cross Abstract: Small language models are usually built like large ones and then squeezed onto a CPU afterwards. We did the opposite: we fixed the target first, one user, one token at a time, 4-bit weights, ordinary CPU, and chose the architecture to suit it. The re...
122. Gravitational-wave parameter estimation with machine-learning generated surrogate waveforms ​
Author: Suyog Garg, Kipp Cannon
Published: 8/21/2026, 4:00:00 AM
Categories: gr-qc, astro-ph.IM, cs.LG
arXiv:2608.20222v1 Announce Type: cross Abstract: The worldwide network of gravitational-wave detectors have detected more than 350 binary coalescence events till date. Future third-generation detectors, like Einstein telescope, are expected to detect orders-of-magnitude more signals from sources wi...
123. Transfer Learning in Nonparametric Regression with Deep ReLU Networks ​
Author: Junpeng Ren, Carlos Misael Madrid Padilla, Yanzhen Chen, Oscar Hernan Madrid Padilla
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2608.20255v1 Announce Type: cross Abstract: This paper develops a general transfer learning framework for nonparametric regression with data consisting of multiple groups. Under the assumption that groups share a common structure along with group-specific deviations in additive form, the propo...
124. Which Eviction Policy Should an LLM Cache Use? A Systematic Study Across Workloads, Capacities, and Encoders ​
Author: Yash Kulkarni, Shubham Harkare, Arvind Suresh Yogesh Babu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.DB, cs.LG
arXiv:2608.20280v1 Announce Type: cross Abstract: Semantic caches reuse an LLM response when the incoming query embedding lies near a cached query, but proposed eviction policies have rarely been compared under one protocol. Using CLEVER, we evaluate FIFO, LRU, LFU, ARC, GDSF, a single-pass streamin...
125. AI4AI-Bench: Benchmarking LLM Agents in Algorithmic Design for Recursive Self-Improvement ​
Author: Yizhe Chi, Wenyi Li, Deyao Hong, Xiaoqiu Wang, Mingju Gao, Kaisen Yang, Bingxiang He, Youjie Zheng, Calvin Xiao, Qinhuai Na
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.20318v1 Announce Type: cross Abstract: Recursive self-improvement (RSI) asks whether an AI system can improve the process that produces AI systems, so that the next system inherits the improvement. That process is the training algorithm: a better objective or update rule improves the comp...
126. $TCP_\alpha$: Margin-Controlled Confidence estimation for reliable Music Information Retrieval ​
Author: Parampreet Singh, Anushka Singh, Sumit Kumar, Vipul Arora
Published: 8/21/2026, 4:00:00 AM
Categories: eess.AS, cs.LG
arXiv:2608.20326v1 Announce Type: cross Abstract: Deep neural networks are often overconfident, assigning high confidence even to incorrect predictions. Consequently, users lack a reliable signal for deciding when a prediction can be trusted. Post-hoc confidence estimation addresses this by training...
127. Information on trajectories: martingales and random times ​
Author: Akshay Balsubramani
Published: 8/21/2026, 4:00:00 AM
Categories: math.PR, cs.IT, cs.LG, math.IT, math.ST, stat.TH
arXiv:2608.20337v1 Announce Type: cross Abstract: Accounting for information flow on the path space of trajectories of a nonnegative martingale yields exact variational identities for it, even at arbitrary random times. This recovers the widely used classical concentration inequalities, from Ville t...
128. On the convergence of optimistic policy iteration for stochastic shortest path problem ​
Author: Yuanlong Chen
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:1808.08763v3 Announce Type: replace Abstract: In this paper, we prove some convergence results of a special case of optimistic policy iteration algorithm for stochastic shortest path problem. We consider both Monte Carlo and $TD(\lambda)$ methods for the policy evaluation step under the condit...
129. Multi-Modal Graph Interaction for Multi-Graph Convolution Network in Urban Spatiotemporal Forecasting ​
Author: Lingyu Zhang, Xu Geng, Zhiwei Qin, Hongjun Wang, Xiao Wang, Ying Zhang, Jian Liang, Guobin Wu, Xuan Song, Yunhai Wang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:1905.11395v2 Announce Type: replace Abstract: Graph convolution network based approaches have been recently used to model region-wise relationships in region-level prediction problems in urban computing. Each relationship represents a kind of spatial dependency, like region-wise distance or fu...
130. Learning-Based Speed Estimation from Accelerometer-Only Inertial Sensing ​
Author: Barak Or
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2401.07468v4 Announce Type: replace Abstract: The proposed model, CarSpeedNet, estimates scalar vehicle speed from a window of three-axis smartphone acceleration, without gyroscope, wheel-odometry, vehicle-bus, or positioning input at inference. The reported experiment comprises 13.2 hours of ...
131. Towards Efficient Pareto Set Approximation via Mixture of Experts Based Model Fusion ​
Author: Anke Tang, Li Shen, Yong Luo, Shiwei Liu, Han Hu, Bo Du, Dacheng Tao
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2406.09770v2 Announce Type: replace Abstract: Solving multi-objective optimization problems for large deep neural networks is a challenging task due to the complexity of the loss landscape and the expensive computational cost of training and evaluating models. Efficient Pareto front approximat...
132. ReAugment: Model Zoo-Guided RL for Few-Shot Time Series Augmentation and Forecasting ​
Author: Haochen Yuan, Yutong Wang, Yihong Chen, Yunbo Wang, Xiaokang Yang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2409.06282v5 Announce Type: replace Abstract: Time series forecasting, particularly in few-shot learning scenarios, is challenging due to the limited availability of high-quality training data. To address this, we present a pilot study on using reinforcement learning (RL) for time series data ...
133. Virtual Sensing to Enable Real-Time Monitoring of Inaccessible Locations & Unmeasurable Parameters ​
Author: Kazuma Kobayashi, Farid Ahmed, Jaewan Park, Subhankar Sarkar, Seid Koric, Souvik Chakraborty, Syed Bahauddin Alam
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP
arXiv:2412.00107v3 Announce Type: replace Abstract: Real-time monitoring of safety-critical interior states is an open problem across energy, environmental and industrial systems where direct instrumentation is infeasible. Approaches based on governing equations, discrete state vectors or fixed sens...
134. Table2Image: Lightweight Tabular Learning with Generated Proxy Representations and Reliability Diagnostics ​
Author: Seungeun Lee, Kihwan Lee, Subin Bae, Sangjun Lee, Seulbin Lee, Julia Stoyanovich, Il-Youp Kwak, Seungsang Oh
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2412.06265v3 Announce Type: replace Abstract: Deep tabular models should ideally balance predictive performance, parameter efficiency, and robustness to imperfect learning signals---properties that are rarely considered jointly. We present Table2Image, a lightweight tabular learning model buil...
135. DeepConvContext: A Multi-Scale Approach to Timeseries Classification in Human Activity Recognition ​
Author: Marius Bock, Juergen Gall, Michael Moeller, Kristof Van Laerhoven
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, eess.IV
arXiv:2505.20894v3 Announce Type: replace Abstract: Despite recognized limitations in modeling long-range temporal dependencies, Human Activity Recognition (HAR) has traditionally relied on a sliding window approach to segment labeled datasets. Deep learning models like the DeepConvLSTM typically cl...
136. ProteinZero: Self-Improving Protein Generation via Online Reinforcement Learning ​
Author: Ziwen Wang, Jiajun Fan, Ruihan Guo, Thao Nguyen, Heng Ji, Ge Liu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2506.07459v4 Announce Type: replace Abstract: Protein generative models have shown remarkable promise in protein design, yet their success rates remain constrained by reliance on curated sequence-structure datasets and by misalignment between supervised objectives and real design goals. We pre...
137. Why Can't I See My Clusters? A Precision-Recall Approach to Dimensionality Reduction Validation ​
Author: Diede P. M. van der Hoorn, Alessio Arleo, Fernando V. Paulovich
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.04222v2 Announce Type: replace Abstract: Dimensionality Reduction (DR) is widely used for visualizing high-dimensional data, often with the goal of revealing expected cluster structure. However, such a structure may not always appear in the projections. Existing DR quality metrics assess ...
138. GraphPFN: A Prior-Data Fitted Graph Foundation Model ​
Author: Dmitry Eremeev, Oleg Platonov, Gleb Bazhenov, Artem Babenko, Liudmila Prokhorenkova
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.21489v4 Announce Type: replace Abstract: Graph foundation models face several fundamental challenges including transferability across diverse domains and data scarcity, which calls into question the very feasibility of creating such models. However, despite similar challenges, the tabular...
139. Merge Now, Regret Later: The Hidden Cost of Model Merging Is Adversarial Transferability ​
Author: Mauro Conti, Ankit Gangwal, Aaryan Ajay Sharma
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.23689v2 Announce Type: replace Abstract: Model Merging (MM) has proven to be an effective alternative to multi-task learning, where several fine-tuned models are merged, without access to the tasks' training data, into one model that retains performance across different tasks. Recent work...
140. R\'enyi Sharpness: A Novel Sharpness that Strongly Correlates with Generalization ​
Author: Qiaozhe Zhang, Jun Sun, Ruijie Zhang, Yingzhuang Liu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.07758v3 Announce Type: replace Abstract: Sharpness (of the loss minima) is widely believed to be a good indicator of generalization of neural networks. Unfortunately, the correlation between existing sharpness measures and generalization is not as strong as expected, and sometimes even co...
141. Towards Formalizing Reinforcement Learning Theory: A Robbins-Siegmund Approach ​
Author: Shangtong Zhang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2511.03618v2 Announce Type: replace Abstract: In this paper, we formalize the almost sure convergence of $Q$-learning and linear temporal difference (TD) learning with Markovian samples using the Lean 4 theorem prover based on the Mathlib library. $Q$-learning and linear TD are among the earli...
142. CarBench: A Comprehensive Benchmark for Neural Surrogates on High-Fidelity 3D Car Aerodynamics ​
Author: Mohamed Elrefaie, Dule Shu, Matt Klenk, Faez Ahmed
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.07847v2 Announce Type: replace Abstract: Benchmarking has been the cornerstone of progress in computer vision, natural language processing, and the broader deep learning domain, driving algorithmic innovation through standardized datasets and reproducible evaluation protocols. The growing...
143. Partition of Unity Neural Networks for Interpretable Classification with Explicit Class Regions ​
Author: Akram Aldroubi
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2602.00511v3 Announce Type: replace Abstract: We introduce \emph{Partition of Unity Neural Networks} (PUNNs), a neural-network architecture for multiclass classification based on the classical mathematical notion of a partition of unity. The starting point is the observation that the character...
144. Maximum Likelihood Reinforcement Learning ​
Author: Fahim Tajwar, Guanning Zeng, Yueer Zhou, Yuda Song, Daman Arora, Yiding Jiang, Jeff Schneider, Ruslan Salakhutdinov, Haiwen Feng, Andrea Zanette
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.02710v2 Announce Type: replace Abstract: Reinforcement learning (RL) is the method of choice for training models in setups where the objective function can only be evaluated by sampling from the model. Our key observation is that when the feedback is terminal and binary, models implicitly...
145. HBVLA: Pushing 1-Bit Post-Training Quantization for Vision-Language-Action Models ​
Author: Xin Yan, Zhenglin Wan, Feiyang Ye, Xingrui Yu, Hangyu Du, Yang You, Ivor Tsang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.13710v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models enable instruction-following embodied control, but their large compute and memory footprints hinder deployment on resource-constrained robots and edge platforms. While reducing weights to 1-bit precision through ...
146. Guided Diffusion by Optimized Loss Functions on Relaxed Parameters for Inverse Material Design ​
Author: Jens U. Kreber, Christian Wei{\ss}enfels, Joerg Stueckler
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.CV
arXiv:2602.15648v2 Announce Type: replace Abstract: Inverse design problems are common in engineering and materials science. The forward direction, i.e., computing output quantities from design parameters, typically requires running a numerical simulation, such as a FEM, as an intermediate step, whi...
147. Neural Prior Estimation: Learning Class Priors from Latent Representations ​
Author: Masoud Yavari, Payman Moallem
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2602.17853v2 Announce Type: replace Abstract: Logit adjustment corrects class imbalance using the empirical class prior. We study whether a comparable class-frequency signal can instead be learned from the network representation, without explicitly supplying class counts to the correction rule...
148. From Rebound to Remedy: Understanding and Mitigating Reward Hacking via Representation Engineering ​
Author: Rui Wu, Ruixiang Tang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2604.01476v3 Announce Type: replace Abstract: Reinforcement learning for LLMs is vulnerable to reward hacking, where models exploit shortcuts to maximize reward without solving the intended task. We systematically study this phenomenon in coding tasks using an environment-manipulation setting,...
149. Transformer See, Transformer Do: Copying as an Intermediate Step in Learning Analogical Reasoning ​
Author: Philipp Hellwig, Willem Zuidema, Claire E. Stevenson, Martha Lewis
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2604.06501v2 Announce Type: replace Abstract: Analogical reasoning is a hallmark of human intelligence, enabling us to solve new problems by transferring knowledge from one situation to another. Yet, developing artificial intelligence systems capable of robust human-like analogical reasoning h...
150. A Robust In-Context Model for Conservation Laws: Injecting Context into Flux Neural Operators via Recurrent Vision Transformers ​
Author: Taeyoung Kim, Joon-Hyuk Ko
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.05488v2 Announce Type: replace Abstract: We propose an architecture that augments the Flux Neural Operator (Flux NO), which combines the classical finite volume method (FVM) with neural operators, with ViT-based context injection. Our model is formulated as a hypernetwork: it extracts sol...
151. Higher Resolution, Better Generalization: Unlocking Visual Scaling in Deep Reinforcement Learning ​
Author: Raphael Trumpp, "Omer Veysel \c{C}a\u{g}atan, Bar{\i}\c{s} Akg"un, Marco Caccamo
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.10546v2 Announce Type: replace Abstract: Pixel-based deep reinforcement learning agents are typically trained on heavily downsampled visual observations, a convention inherited from early benchmarks rather than grounded in principled design. In this work, we show that observation resoluti...
152. RequestRouter: Request-Boundary Routing for Efficient Single-GPU LLM Inference ​
Author: Aman Sunesh, Ali Alshehhi, Hivansh Dhakne
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.PF
arXiv:2605.23057v2 Announce Type: replace Abstract: RequestRouter is a lightweight request-boundary controller for reducing the latency and energy cost of single-GPU large language model inference. Rather than serving all requests with one static configuration, the system uses cheap request-level fe...
153. The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth ​
Author: James Henry
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.24856v2 Announce Type: replace Abstract: Concept formation in transformer language models is a depth-extended process, not a single-layer event: a concept becomes separable across one or more contiguous regions of the residual stream - its Concept Allocation Zone (CAZ). A CAZ is not a con...
154. Geometric Evolution Maps: Extracting Stable Concept Probes from Transformer Residual Streams ​
Author: James Henry
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.25848v2 Announce Type: replace Abstract: A concept probe is only as reliable as the layer it is taken from. Probing at a fixed late layer, or at the peak of a separation curve, ignores a structural feature of how concepts form: the probe direction rotates substantially during assembly and...
155. The Good, the Bad, and the Ugly of Markov Boundary for Tabular Prediction ​
Author: Shu Wan, Abhinav Gorantla, Huan Liu, K. Sel\c{c}uk Candan
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ME, stat.ML
arXiv:2605.29411v2 Announce Type: replace Abstract: Under standard graphical assumptions, the Markov boundary of a target variable is the smallest set of features that renders every other feature redundant. Once the boundary is observed, the target is conditionally independent of the rest of the tab...
156. GEM: Geometric Erasure by Contrastive Velocity Matching in Rectified Flows ​
Author: Jonas Henry Grebe, Tobias Braun, Anna Rohrbach, Marcus Rohrbach
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.00140v3 Announce Type: replace Abstract: While the rapid adoption of multimodal generative models offers immense potential, it has also increased the risks of harmful content synthesis, deepfakes, and copyright infringements. To address these challenges, concept erasure has emerged as a p...
157. From Prediction to Self: Developmental Conditions for Agency in Minimal Neural Systems ​
Author: Evan Ye
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2606.05605v2 Announce Type: replace Abstract: How does a system that merely predicts the world come to distinguish its own causal influence from everything else? We trace this transition in a minimal 192-dimensional GRU through a developmental sequence -- 6 experimental stages, 12 falsified al...
158. Bootstrap Theory of Representational Emergence (TBER): Explanatory Insufficiency, Transition Regimes, and the Emergence of New Representational Levels ​
Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.07303v4 Announce Type: replace Abstract: Representation learning is central to modern machine learning, yet most research focuses on optimizing representations after a framework has been selected. The Bootstrap Theory of Representational Emergence (TBER) addresses a prior question: when d...
159. PROBE-Web: An Interactive System for Probing Evaluation Landscapes of Knowledge Graph Completion Models ​
Author: Sooho Moon, Yunyong Ko
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.08926v2 Announce Type: replace Abstract: Knowledge graph completion (KGC) models are commonly evaluated using rank-based metrics such as MRR and Hits@K, despite different users often requiring different evaluation perspectives. In this demo, we present PROBE-Web, an interactive system for...
160. Detecting Explanatory Insufficiency in Learned Representations: A Framework for Representational Vigilance ​
Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.13172v4 Announce Type: replace Abstract: Learned representations are central to modern machine learning, but predictive performance, robustness, uncertainty estimation, and generalization do not by themselves establish representational adequacy. A model may remain operationally successful...
161. Generalist Vision-Language Models for Fast Radio Burst detection: a zero-shot benchmark against a specialized detector ​
Author: Raiff H. Santos, Amilcar R. Queiroz, Tharcisyo S. S. Duarte, K. E. L. de Farias, Rafael A. Batista
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.HE, astro-ph.IM
arXiv:2607.07382v2 Announce Type: replace Abstract: Fast Radio Burst (FRB) detection increasingly relies on specialized deep learning models that require large task-specific training sets and cannot be redefined without retraining. We evaluate whether small, open-weight, locally run generalist Visio...
162. Grounded verification of chemical and materials reasoning: detection is the bottleneck ​
Author: Can Polat, Mustafa Kurban, Erchin Serpedin, Hasan Kurban
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph, physics.comp-ph, quant-ph
arXiv:2607.17417v2 Announce Type: replace Abstract: Language models are moving into chemistry and materials discovery workflows, where a wrong molecular formula, space group, or formation energy can silently propagate into downstream decisions. These confabulations hide inside fluent reasoning trace...
163. Drift-Adaptive ICU Intervention Prediction: Freezing the Physiological Encoder for Auditable Model Updating ​
Author: Fatema Ferdous Tamanna, K. M. Merajul Arefin, Md. Abdul Masud
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR, q-bio.QM
arXiv:2607.19020v2 Announce Type: replace Abstract: Clinical decision support degrades as treatment protocols evolve, but the obstacle to updating a deployed model is governance as much as accuracy: once retraining touches every parameter, no one can say afterwards where the update acted. We propose...
164. DASH-OPD: Discrepancy-Aware Switching with Hysteresis for On-Policy Distillation ​
Author: Yuchen Xia, Qianguo Sun, Chao Song, Junlong Wu, Yiyan Qi, Yunjian Xu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29078v2 Announce Type: replace Abstract: On-policy distillation (OPD) trains student models on their own rollouts to reduce exposure bias. However, in multi-turn agent scenarios, early student errors can lead a trajectory away from the teacher's familiar domain. Existing curriculum learni...
165. A Neurosymbolic Approach for Explainable Early Diagnosis of Alzheimer's Disease ​
Author: Ranveer Singh, Pranuthi Tenali, Saurabh Mathur, Ameet Soni, Vaishali Phatak, Karla Lynch, Daniel Murman, Matthew Rizzo, Sriraam Natarajan
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.29530v2 Announce Type: replace Abstract: Identifying reliable Alzheimer's disease (AD) markers typically requires manual, labor-intensive transcription and expert analysis, limiting its scale. We introduce an automated pipeline that extracts qualitative knowledge about potential AD progre...
166. FinVerse: Financial Time-Series Benchmark ​
Author: Jaehoon Lee, Jun Seo, Seunghan Lee, Tae Yoon Lim, Dongwan Kang, Hwanil Choi, Minjae Kim, Sungdong Yoo, Junhyeok Kang, Sangjun Han, Soonyoung Lee, Wonbin Ahn
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.03259v2 Announce Type: replace Abstract: As time-series foundation models have emerged, the need for benchmarks that can evaluate their forecasting ability in meaningful ways has become increasingly important. Existing time-series forecasting benchmarks provide useful standardized compari...
167. Graph Machine: Exploring Edge Mechanisms as an Inductive Bias ​
Author: Lintai Hou
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06834v2 Announce Type: replace Abstract: Transformers provide a powerful architecture for global content-based matching, but reasoning problems may benefit from a stronger inductive bias toward iterative traversal of latent relations. We introduce Graph Machine, an architecture with two e...
168. Why AI Detection Fails for Academic Integrity ​
Author: Jonathan A. Karr Jr, Grigorii Khvatskii, Ting Hua, Nitesh V. Chawla
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CY
arXiv:2608.11256v2 Announce Type: replace Abstract: Institutions use commercial AI detectors for academic integrity, yet detectors cannot distinguish AI editing from full LLM drafts and may treat both as misconduct. In a controlled study of published English abstracts (four domains; 2013 to 2015 vs....
169. Geometric and Behavioral Stratification in Transformer Residual Streams ​
Author: Nelson Guda
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.12447v3 Announce Type: replace Abstract: Trained transformer models develop privileged bases: coordinate axes whose statistics differ from the rest of the residual stream. But what kind of direction does such a basis select? We investigate the prediction direction, the unembedding directi...
170. Agents unlock new capabilities through Switching LoRA Adapters as a Tool (SLAaaT) ​
Author: Kenneth Ge
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17034v2 Announce Type: replace Abstract: Post-training can unlock new capabilities and improve performance on specialized tasks, but sometimes at the cost of catastrophic forgetting in other domains. This poses a problem in long agent trajectories that compose different capabilities. We r...
171. Too Sure to Be Safe: Model Calibration for Reliable Log Anomaly Detection ​
Author: Bin Li, Dongdong Wang, Siyang Lu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE
arXiv:2608.17965v2 Announce Type: replace Abstract: Online log anomaly detection is critical for maintaining the reliability of large-scale computing systems. Although recent language model-based log anomaly detectors achieve strong detection performance, their confidence estimates remain poorly cal...
172. Continual Reasoning Gym: Diagnosing and Harnessing Shared Reasoning in Continual RLVR ​
Author: Lirui Luo, Guoxi Zhang, Hongming Xu, Rongqing Li, Cong Fang, Lifeng Fan
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.18574v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) commonly post-trains reasoning models on multiple tasks, while rerunning multitask RLVR (MTRL) as new tasks are added makes capability expansion costly. We therefore study continual RLVR, which ...
173. Graphical Design of Interpretable Architectures ​
Author: Pietro Barbiero
Published: 8/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2608.18936v2 Announce Type: replace Abstract: Designing, implementing, and comparing interpretable architectures requires a formal language to represent them. The most common representations fall short in one of two ways. Symbolic equations give no global view of an architecture at a glance. P...
174. Asymptotic Theory for IV-Based Reinforcement Learning with Potential Endogeneity ​
Author: Jin Li, Ye Luo, Zigan Wang, Xiaowei Zhang
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, econ.EM, math.OC
arXiv:2103.04021v4 Announce Type: replace-cross Abstract: In the standard data analysis framework, data is collected (once and for all), and then data analysis is carried out. However, with the advancement of digital technology, decision-makers constantly analyze past data and generate new data thro...
175. Teacher-free Latent Self-distillation and Class-separable Representations for Lightweight IoT Attack Detection ​
Author: Phai Vu Dinh, Diep N. Nguyen, Dinh Thai Hoang, Marwan Krunz, Quang Uy Nguyen, Son Pham Bao, Eryk Dutkiewicz
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2403.15509v3 Announce Type: replace-cross Abstract: Knowledge distillation (KD) has been widely used to improve lightweight AI models by transferring soft-label knowledge from a large teacher model to a student model. However, existing KD methods are primarily designed for the image domain rat...
176. Decoupling High and Low Frequencies for Faithful Image Generation with Fine Details ​
Author: Tejaswini Medi, Hsien-Yi Wang, Arianna Rampini, Margret Keuper
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2509.05441v4 Announce Type: replace-cross Abstract: Latent generative models compress images into learned embeddings prior to synthesis, and the generation quality critically depends on how faithfully these embeddings preserve visual detail. We observe that while such embeddings are effective ...
177. MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal Prostate MRI Segmentation ​
Author: Yovin Yahathugoda, Davide Prezzi, Patricia A. Gutierrez, Piyalitt Ittichaiwong, Vicky Goh, Sebastien Ourselin, Michela Antonelli
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2510.17529v3 Announce Type: replace-cross Abstract: Active Surveillance (AS) is a treatment option for managing low and intermediate-risk prostate cancer (PCa), aiming to avoid overtreatment while monitoring disease progression through serial MRI and clinical follow-up. Accurate prostate segme...
178. Better Call Graphs: A New Dataset of Function Call Graphs for Malware Classification ​
Author: Jakir Hossain, Jue Guo, Gurvinder Singh, Lukasz Ziarek, Ahmet Erdem Sar{\i}y"uce
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2512.20872v2 Announce Type: replace-cross Abstract: Function call graphs (FCGs) have emerged as a powerful abstraction for malware detection, capturing the behavioral structure of applications beyond surface-level signatures. Their utility in traditional program analysis has been well establis...
179. Reducing the Complexity of Matrix Multiplication by Quantum Computing ​
Author: Jiaqi Yao, Tianjian Huang, Tonghe Zhang, Ding Liu
Published: 8/21/2026, 4:00:00 AM
Categories: quant-ph, cs.CC, cs.LG
arXiv:2602.05541v3 Announce Type: replace-cross Abstract: Matrix multiplication is a fundamental operation in compute-intensive tasks and a key component of modern quantum acceleration frameworks. Here we present a quantum matrix multiplication algorithm based on quantum kernels (QKMM), achieving an...
180. A Unified Physics-Informed Neural Network for Modeling Coupled Electro- and Elastodynamic Wave Propagation Using Three-Stage Loss Optimization ​
Author: Suhas Suresh Bharadwaj, Reuben Thomas Thovelil
Published: 8/21/2026, 4:00:00 AM
Categories: cs.NE, cs.LG, physics.comp-ph
arXiv:2602.13811v3 Announce Type: replace-cross Abstract: Physics-Informed Neural Networks present a novel approach in SciML that integrates physical laws in the form of partial differential equations directly into the NN through soft constraints in the loss function. This work studies the applicati...
181. Learn for Variation: Efficient AAV Trajectory Learning through a Differentiable Wireless World Model ​
Author: Xiucheng Wang, Zhenye Chen, Nan Cheng, Zhisheng Yin, Xuemin Shen
Published: 8/21/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2603.18853v3 Announce Type: replace-cross Abstract: Autonomous aerial vehicles (AAVs) enable data collection for sixth-generation Internet-of-Things networks, but their trajectories couple nonlinear wireless rates with long-horizon service progress. This paper views the evolution of AAV kinema...
182. The Value of Finite Observation in Positive-Data Learning of Multiple Context-Free Languages ​
Author: Takayuki Kuriyama
Published: 8/21/2026, 4:00:00 AM
Categories: cs.FL, cs.LG
arXiv:2605.11644v3 Announce Type: replace-cross Abstract: Positive data can show that two tuple occurrences share a successful sentence context without certifying that they are safely interchangeable. We study finite compositional observations, represented by finite-monoid homomorphisms, as semantic...
183. Diffusion-based Denoising Beats Vanilla Score Matching in Parameter Estimation: A Theoretical Explanation ​
Author: Benedikt L"utke Schwienhorst, Nadja Klein, Johannes Lederer
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.ME, stat.TH
arXiv:2605.22950v2 Announce Type: replace-cross Abstract: Score matching is an alternative to maximum likelihood estimation when the normalizing constant is unknown or too costly to evaluate. However, vanilla score matching has shown to be inefficient relative to maximum likelihood estimation for mu...
184. From Accuracy to Auditability: A Survey of Determinism in Financial AI Systems ​
Author: Ruizhe Zhou, Xiaoyang Liu, Gaoyuan Du, Yi Zheng, Shouxi Ren, Deepayan Chakrabarti, Dengdu Jiang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.DC, cs.LG, cs.SI, q-fin.CP
arXiv:2605.23955v3 Announce Type: replace-cross Abstract: Deploying machine learning in regulated financial environments -- credit risk, fraud detection, and anti-money laundering -- exposes critical vulnerabilities in algorithmic reproducibility. While early financial ML addressed statistical chall...
185. Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization ​
Author: Jingyi Sun, Qianli Wang, Pepa Atanasova, Nils Feldhus, Isabelle Augenstein
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2605.24960v2 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) faithfulness, i.e., whether CoTs genuinely reflect large language models' (LLM) underlying behavior, is typically evaluated with metrics under two disjoint paradigms: contextual faithfulness, measured by perturbing the ...
186. Continuous Behavioral Authentication via Multi-Expert BERT Log Analysis for Secure Data Sharing ​
Author: Stergios Lantzos, Ilias Syrigos, Apostolos Apostolaras, Thanasis Korakis
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2606.21900v2 Announce Type: replace-cross Abstract: Continuous authentication for mobile and zero-trust systems requires nonintrusive evidence confirming the enrolled user-device context remains valid after initial login. This paper presents a BERT log analysis framework for continuous behavio...
187. MulTTiPop: A Multitrack Transcription Dataset for Pop Music ​
Author: Nathan Pruyne, Benjamin Stoler, William Chen, Chien-yu Huang, Shinji Watanabe, Chris Donahue
Published: 8/21/2026, 4:00:00 AM
Categories: cs.SD, cs.LG
arXiv:2607.08756v2 Announce Type: replace-cross Abstract: We present MulTTiPop, a dataset of pop music segments and their associated multitrack MIDI recordings for the evaluation of automatic music transcription models. MulTTiPop contains 572 segments of popular music totaling 3.5 hours of audio, an...
188. The Price of Hidden Curvature: Improved Lower Bounds for Bandit Convex Optimization ​
Author: Nived Rajaraman, Yanjun Han
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT
arXiv:2607.18652v3 Announce Type: replace-cross Abstract: We establish improved lower bounds on the minimax expected regret of stochastic bandit convex optimization for $1$-Lipschitz functions on the $d$-dimensional Euclidean ball. For time horizons $n\ge d^{10/3}$, we prove a lower bound of $\Omega...
189. Distributional Determinantal Point Process for Repulsive Clustering of Distributions ​
Author: Khai Nguyen, Yang Ni, Elizabeth Juarez-Colunga, Peter Mueller
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.AP, stat.CO, stat.ML
arXiv:2607.21847v2 Announce Type: replace-cross Abstract: We introduce the distributional determinantal point process (dDPP) as a novel repulsive point process whose atoms are probability distributions rather than points in a real space. The dDPP is constructed via an L-ensemble with a sliced Wasser...
190. Learning Asymptotics with Convergence-Rate Guarantees using Linear Least Squares ​
Author: Christos N. Efrem
Published: 8/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.CO, math.NA
arXiv:2607.23287v2 Announce Type: replace-cross Abstract: We introduce a new research area that is called Asymptotics Learning Theory (ALT) and combines optimization with asymptotic analysis. In particular, ALT provides a unified approach for computing unknown constants/parameters in proven asymptot...
191. LLM Capability Limits: Static Emergence and Dynamic Boundary Control ​
Author: Yi Liu
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.01548v3 Announce Type: replace-cross Abstract: Test-time emergence in LLM systems has a deployment boundary: additional computation can realize decisions already supported by the deployed information--execution structure, while evidence, tools, memory, and executable semantics can change ...
192. Provably Efficient Self-Calibrating Quantum Fault Tolerance ​
Author: Weiyuan Gong, Hong-Ye Hu
Published: 8/21/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.05686v2 Announce Type: replace-cross Abstract: Quantum error correction protects logical information only when every physical operation remains below the fault-tolerance threshold, a condition that must be maintained continuously rather than only at the initial calibration. In practice, h...
193. Verifiably grounded machine interpretation of lunar geology ​
Author: Tom Sander, Kay Wohlfarth, Christian W"ohler
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.09276v2 Announce Type: replace-cross Abstract: Planetary geology relies on historical, interpretive reasoning to reconstruct past events from diverse observations. Here, we investigate how far this interpretive workflow can be automated by a multimodal vision-language model. Focusing on t...
194. Uncertainty-Aware Compositional Localization and Placement Assessment of Catheters and Tubes in Chest X-Rays ​
Author: Harshil Lodhiya
Published: 8/21/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2608.11288v2 Announce Type: replace-cross Abstract: Assessing catheter and tube placement on chest X-rays is safety-critical yet tedious and error-prone. Current deep learning methods either classify placement globally -- losing track of which device is where -- or segment all devices into a s...
195. LODESTAR: Robust Entropy-Based Answer Selection in Retrieval-Augmented Generation for Question Answering -- Directing Frozen-LLM Entropy with a Reinforcement-Learned Prompt Polarizer under Misleading Passages ​
Author: Hung-Chun Hsu, Po-Jen Ko, Che-Cheng Wu, Li-Yang Chang, Chuan-Ju Wang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.IR, cs.LG
arXiv:2608.11922v2 Announce Type: replace-cross Abstract: Predictive-distribution entropy is a strong answer-selection rule in retrieval-augmented generation (RAG) for question answering: across five QA benchmarks, selecting the answer a frozen respondent LLM produces with the lowest answer-token en...
196. When Is Shallow Enough? Adaptive Split Federated Learning with Client-Specific Sufficiency Estimation ​
Author: Wenhao Yuan, Chenchen Lin, Wentao Hu, Jian Chen, Jinfeng Xu, Shujie Li, Edith Cheuk Han Ngai
Published: 8/21/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG
arXiv:2608.15639v2 Announce Type: replace-cross Abstract: \textit{Split Federated Learning} (SFL) enables distributed model training by splitting networks between the server and clients. However, under client heterogeneity, the conventional static split strategy may be suboptimal because clients can...
197. Ask to Be Sure: Informative Interactions for Confident Multi-Turn LLM Recommendation ​
Author: Cedar Site Bai, Zhenyu Liao, Duanshun Li, Sheikh Sarwar, Huiyuan Chen, Yuan Chen, Changhe Yuan, Haiyang Zhang, Qilin Qi
Published: 8/21/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG
arXiv:2608.15949v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have enabled their use as conversational recommender systems (CRS), demonstrating strong recommendation accuracy and natural dialogue. However, guiding multi-turn interactions to elicit user pre...
198. Automating Learner Assessment: Benchmarking Machine Learning and Deep Learning Models for EEG-Based Familiarity Prediction ​
Author: Isuru Nanayakkara, Thilina Halloluwa
Published: 8/21/2026, 4:00:00 AM
Categories: eess.SP, cs.HC, cs.LG
arXiv:2608.16541v2 Announce Type: replace-cross Abstract: Objective assessment of learning remains a fundamental challenge in education. Electroencephalography (EEG) provides a direct, non-invasive window into the neural correlates of knowledge acquisition, including cognitive familiarity. This stud...
199. Which Negatives Matter? Ask Your Text Encoder: Adaptive Similarity Margins for Dense-Caption Retrieval ​
Author: Haoyue Liu, Ye Chen, Zhichao Wang, Xiaoying Tang
Published: 8/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.18521v2 Announce Type: replace-cross Abstract: Dense-caption retrieval has recently been improved by introducing segmentation, edge maps, LLM-filtered captions, and cross-modal modules into contrastive fine-tuning. However, these methods largely inherit the same InfoNCE objective, whose o...
200. Compress and Forget: bitsandbytes Quantization Amplifies Proactive Interference in LLMs ​
Author: Shayan Shahrabi-Farahani, Dara Rahmati
Published: 8/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.18578v2 Announce Type: replace-cross Abstract: Proactive interference (PI) is a documented failure mode in large language models in which retrieval of a repeatedly overwritten value degrades as prior overwrites accumulate, mirroring a classical phenomenon in human working memory. Post-tra...