arXiv cs.LG - 2026-07-22 ​
284 items collected.
1. FALCON-Discover: Discovering Concentrated False-Confidence Regions for Calibration ​
Author: Filippo Cenacchi, Longbing Cao, Runze Yang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18278v1 Announce Type: new Abstract: Calibration is usually evaluated in aggregate, but the most dangerous failures are often local: predictions that remain highly confident despite being wrong. We study this failure mode as false-confidence concentration, the extent to which confident er...
2. Beyond Output-Space Calibration: Spectral Evidence Bundling for Selective Reliability Estimation in Time-Series Classification ​
Author: Filippo Cenacchi, Longbing Cao, Runze Yang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18279v1 Announce Type: new Abstract: Post-hoc calibration for time-series classification usually remaps output scores, but deployment decisions such as trust, abstention, and review depend on whether a confident prediction is supported by the current temporal signal. We address three time...
3. Beyond Single-Dimensional Compression: The Compound Sparsity Frontier of Large Language Models ​
Author: Chao Han, Haozhe Hu, Xiaoyu Shen
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18280v1 Announce Type: new Abstract: Large language models (LLMs) are often compressed through static parameter pruning or dynamic token-level computation, yet aggressive sparsification can trigger rapid performance degradation beyond an essential sparsity boundary. This work asks \emph{w...
4. ALAS: Additive Learnable Alpha-Stable Kernels for Flexible Bayesian Optimization ​
Author: Weibo Huang, Cheng Hua
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2607.18282v1 Announce Type: new Abstract: Bayesian Optimization is widely used for expensive black-box optimization, yet its success often depends on choosing a kernel that matches the objective's unknown structure. In this work, we propose ALAS, a flexible Gaussian Process kernel family built...
5. FedCC: A Low-Resource Federated Adaptation of Foundation Models for Robust Corpus Callosum localization in Fetal Ultrasound Images ​
Author: Alessandro Di Matteo, Sara Moccia, Giuseppe Rizzo, Gianpaolo Grisolia, Ricciarda Raffaelli, Lorenzo Vasciaveo, Francesco D'Antonio, Maria Chiara Fiorentino
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2607.18283v1 Announce Type: new Abstract: Accurate localization of the corpus callosum (CC) in fetal ultrasound (US) images is crucial for the early identification of neurodevelopmental abnormalities. However, this task remains highly challenging due to the intrinsic limitations of US imaging,...
6. Compressing What Matters: Neuron Importance Meets Data-Aware Low Rank Approximation for Language Model Compression ​
Author: Athanasios Ntovas, Alexandros Doumanoglou, Petros Drakoulis, Dimitris Zarpalas
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18284v1 Announce Type: new Abstract: To excel at their domain large language models are comprised of billions of parameters. Yet this comes at the cost of huge memory requirements restricting their applicability in resource-constrained environments. To address the problem of neural networ...
7. Edge-Efficient Transformer for End-to-End RF Spectrum Monitoring ​
Author: Zhifan Song, Haralampos-G. Stratigopoulos, Hassan Aboushady
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18285v1 Announce Type: new Abstract: We present E-SpecFormer (Edge Spectrum monitoring Transformer) for end-to-end automatic modulation and covert channel (CC) recognition. We introduce LiTAN (Linear Tanh Attention Network), a Softmax- and LayerNorm-free attention mechanism that reduces c...
8. Preference-Conditioned Multi-Objective Reinforcement Learning for Runtime-Tunable Transit Signal Priority ​
Author: Philip-Roman Adam, Stefanie Schmidtner
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY
arXiv:2607.18286v1 Announce Type: new Abstract: Transit signal priority (TSP) requires balancing competing objectives: reducing bus delay while limiting adverse impacts on non-bus traffic and avoiding extreme waits for a subset of vehicles. Existing reinforcement-learning (RL) approaches to TSP typi...
9. BearingNAS: Obtaining In-Sensor Intelligent Fault Diagnosis Systems for Bearings Using a Laptop ​
Author: Andrea Mattia Garavagno, Edoardo Ragusa, Paolo Gastaldo, Antonio Frisoli, Rodolfo Zunino
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18287v1 Announce Type: new Abstract: This paper introduces BearingNAS, a Hardware-Aware Neural Architecture Search (HW-NAS) framework designed to shift the intelligence directly onto the sensor die via in-sensor processing. BearingNAS frames the search as a constrained optimization proble...
10. Multi-Timescale Latent-Action DRL for Joint Optimization in Edge-Cloud Networks ​
Author: Vo Phi Son, Van-Dinh Nguyen, Ngoc Hung Nguyen, Trinh Van Chien, Symeon Chatzinotas
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18288v1 Announce Type: new Abstract: Load imbalance across edge and cloud layers degrades latency performance in hierarchical edge-cloud computing (HECC) systems under dynamic task arrivals and heterogeneous resources, leading to severe queuing delays and inefficient resource utilization....
11. Towards Principled Continual Anomaly Detection: A Systematic Framework and Benchmark Scenarios ​
Author: Kamil Faber, Mateusz Smendowski, Roberto Corizzo
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18289v1 Announce Type: new Abstract: Continual anomaly detection (CAD) studies how models can adapt to evolving data distributions while retaining performance on previously observed regimes. CAD benchmarks, however, depend critically on how tasks are defined, filtered, ordered, and valida...
12. SechKAN: Kolmogorov-Arnold Networks with Hyperbolic Secant Functions ​
Author: Hoang-Thang Ta
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18290v1 Announce Type: new Abstract: In recent years, Kolmogorov-Arnold Networks (KANs) have attracted increasing attention due to their effectiveness in machine learning and scientific computing tasks, offering a new paradigm for neural network design. In this paper, we present SechKAN, ...
13. Dual-domain fused LSTM modeling for efficient time-dependent reliability analysis ​
Author: Yixin Zhang, Mingyang Li, Zichao Jiang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE
arXiv:2607.18291v1 Announce Type: new Abstract: Time-dependent reliability analysis is crucial for ensuring the long-term safety and performance of engineering systems under uncertainties. However, traditional surrogate model methods often struggle to incorporate time-independent random variables an...
14. Reliability Scales Inversely: Bigger Models Compound Mistakes Faster via a Hidden Auto-Regressive Risk Regime ​
Author: Kushal Chakrabarti
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.18292v1 Announce Type: new Abstract: As language models scale, answers start truer but degrade faster: scaling buys capability but erodes reliability. The knowledge-gap account - more data, retrieval, or scale - misses an auto-regressive risk residual that scale sharpens: the model commit...
15. One Student, Many Teachers: Multi-Task On-Policy Distillation via Soft-Prompt Privileged Context ​
Author: Yingzi Ma, Zichen Zhu, Ming Jiang, Chaowei Xiao
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.18293v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) teaches large language models new skills through a teacher that shares the student's backbone and supervises its own rollouts. Existing teachers either inject privileged context at the input -- inducing post-hoc ratio...
16. Uncertainty Quantification for AI-Driven Crash Simulation Surrogates: A Comparative Study of Monte Carlo Dropout and Deep Ensemble on Open-Source Bumper Beam Benchmark ​
Author: Sudeep Chavare
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18294v1 Announce Type: new Abstract: Machine learning surrogate models are increasingly being explored in engineering product development to augment simulation-driven design, offering near-instantaneous predictions that complement computationally expensive high-fidelity analyses. However,...
17. On the Limits of Support-Preserving Alignment and Bounded Filtering ​
Author: Aryan Dutt, Rui Mao, Anupam Chattopadhyay
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18295v1 Announce Type: new Abstract: We study whether alignment schemes that reshape a base model's output distribution, combined with bounded safety filters, can drive the probability of harmful behavior to zero in modern large language models. Recent research suggests that harmful behav...
18. A Better Start for Language Models: Domain-Conditional Position Offsets ​
Author: Ye Qiao
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18302v1 Announce Type: new Abstract: Autoregressive language models are least accurate at the beginning of a sequence, where little context forces reliance on a generic pretraining prior. We show that this cold-start penalty is domain dependent and reduce it with a domain-conditional posi...
19. TD-DPO: Difference-Aware Preference Optimization for Mitigating Sycophancy in Clinical Autism Intervention Dialogue ​
Author: Shuzhong Lai, Junhong Lai, Chenxi Li, Qing Zhou, Haifeng Li, Gang Pan, Lin Yao, Yueming Wang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18304v1 Announce Type: new Abstract: The sycophancy of large language models can increase the safety risk in intervention dialogue for autistic children. Supervised fine-tuning can somewhat reduce sycophancy, but relying solely on positive examples is often insufficient to identify and co...
20. The Information Shadow: Measuring Structural Limits on What Language Models Can Learn ​
Author: Priyansh Srivastava, Romit Chatterjee
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18305v1 Announce Type: new Abstract: Some limits on what language models know are not gaps in data coverage but structural properties of learning from text. We introduce the information shadow: the region of phenomena that a text-trained learner cannot acquire regardless of scale, compris...
21. Gradient-Energy Guided Block-Wise Perturbations for Sharpness-Aware Minimization ​
Author: Zhen Huang, Jiaxin Deng, Junbiao Pang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18306v1 Announce Type: new Abstract: Sharpness-Aware Minimization (SAM) improves generalization by minimizing the worst-case loss in a local parameter neighborhood. Standard SAM implicitly allocates its global perturbation budget across parameter blocks according to instantaneous minibatc...
22. Agentic Calibration of Grey-Box Simulation Models: An LLM-Driven Alternative ​
Author: David G'omez-Guill'en, Mireia Diaz, Josep Lluis Arcos, Jes'us Cerquides
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18308v1 Announce Type: new Abstract: Calibration of grey-box simulation models is a constrained optimization problem in which model evaluations are expensive, the parameter space can be high-dimensional, and the search must respect plausibility constraints. Although the simulation code is...
23. Spatio-Temporal Prediction of Unsteady Airfoil Aerodynamics Using Augmented Graph Neural Ordinary Differential Equations with Exogenous Controls ​
Author: Henrik Lange, Reik Thormann, Philipp Bekemeyer
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn
arXiv:2607.18309v1 Announce Type: new Abstract: Unsteady aerodynamic phenomena, such as gusts, turbulence, and fluid-structure interactions affect an aircraft during flight. For design, optimisation and certification, it is indispensable to quantify such unsteady aerodynamic effects. Industry-standa...
24. Interactive Training 2: Auditable Control Plane for Live Model Training ​
Author: Wentao Zhang, Xuanhe Pan, Han Zhou, Yang Lu, Yuntian Deng
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18314v1 Announce Type: new Abstract: Experiment trackers show how training is progressing, but changing a live run still usually requires trainer-specific code. We present Interactive Training 2, an open-source control plane for steering training through a shared protocol. Training applic...
25. Cost Accounting for Reactive Computational Graphs: Exhaustive Sweeps, Sequential Mutation, and the Backward-Locality Gap ​
Author: Abdallah Khemais (ISITCOM, University of Sousse)
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18323v1 Announce Type: new Abstract: Exhaustive site-by-site interventions on a neural network's computational graph -- activation-patching sweeps, circuit-discovery searches, systematic ablation studies -- mutate the graph at every candidate site, and their cost is dominated by recomputa...
26. Dynamic Loss Balancing for Joint SOH and RUL Prediction of Lithium-Ion Batteries via a Rotary SOH-Injected Prior Battery Transformer ​
Author: Shuhao Chen, Tianyu Shi, Yiwen Huang, Chengyi Tu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18329v1 Announce Type: new Abstract: The deployment of reliable lithium-ion battery management systems is crucial for accelerating electrification, yet the joint prognosis of State of Health (SOH) and Remaining Useful Life (RUL) remains severely hindered by task heteroscedasticity. Conven...
27. Physics-Guided Masked Multi-Task Network for Edge-Friendly Battery Health Diagnostics from Sto-chastically Fragmented Charging Profiles ​
Author: Shuhao Chen, Tianyu Shi, Chengyi Tu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18330v1 Announce Type: new Abstract: The deployment of reliable lithium-ion battery management systems is crucial for accelerating electrification, yet the joint prognosis of State of Health (SOH) and Remaining Useful Life (RUL) remains severely hindered by task heteroscedasticity. Conven...
28. ChemHyperMag: Physics-informed magnetic hypergraph learning improves molecular ADMET prediction ​
Author: Hexiao Ding, Hongzhao Chen, Jing Lan, Yufeng Jiang, Zihong Luo, Zehua Xiong, Tianlong Ruan, Yunlin Mao, Nga Chun Ng, Gwing Kei Yip, Gerald W. Y. Cheng, Kate Inyoung Oh, Jing Cai, Liang-Ting Lin, Jung Sun Yoo
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18332v2 Announce Type: new Abstract: Accurate prediction of ADMET (Absorption, Distribution, Metabolism, Excretion, and Toxicity) is important for drug discovery. Most predictors use undirected molecular graphs and pairwise edges. This choice misses asymmetric interactions, nonreversible ...
29. Federated Lightweight Fine-Tuning ​
Author: Radhakrishna Achanta, Will Reed
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18343v1 Announce Type: new Abstract: Federated fine-tuning is bottlenecked by communication: FedAvg and pseudo-gradient schemes transmit a payload that scales with the model, and gradient compression shrinks it by only a constant factor. We take a different lever. Mapping networks generat...
30. An Analysis of Residual-Stream Geometry Across Transformer Depth ​
Author: Sunit Bhattacharya, Ravi Shankar Kolli
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18348v1 Announce Type: new Abstract: We propose a transition-centred geometric analysis of transformer residual streams. Relative displacement measures how \emph{far} representations move between consecutive layers, and orthogonal Procrustes analysis separates each transition into a rigid...
31. MambaLSTM: A Spatio-Temporal Framework for Enhanced Traffic Accident Risk Prediction ​
Author: Zhen Yu, Yachao Yuan, Zixiang Peng, Muting Li, Thar Baker
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18353v1 Announce Type: new Abstract: In traffic accident risk prediction, most studies overlook the extra noise that could be incorporated when fusing temporal features into spatial features, and some models struggle to capture global correlations among spatial regions. To address these c...
32. Multi-layer MIMO Relay as Deep Physical Neural Networks: Power Amplifiers as Activation Functions ​
Author: Meng Hua, Itsik Bergel, Deniz G"und"uz
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, eess.SP, math.IT
arXiv:2607.18354v1 Announce Type: new Abstract: Wireless physical neural networks (WPNNs) embed neural computation directly into analog hardware, offering lower energy consumption and latency than conventional digital implementations. In this paper, we propose a deep WPNN in which nonlinear activati...
33. Physical Self-Supervised Learning: IMU Sensing without Manual Labels ​
Author: Yuyang Leng (Richard), Renyuan Liu (Richard), Shaohan Hu (Richard), Peijun Zhao (Richard), Chun-Fu Chen (Richard), Songqing Chen, Shuochao Yao
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18361v1 Announce Type: new Abstract: Deep neural networks have become a promising approach for IMU-based sensing, but their scalability is fundamentally limited by costly labeled data and poor robustness to heterogeneous devices, placements, and users. Existing unsupervised and self-super...
34. A Controlled Study of Attention-Only Transformers ​
Author: Henry Ndubuaku, Karen Mosoyan, Jakub Mroz, Noah Cylich, Satyajit Kumar, Parkirat Sandhu, Roman Shemet, Justin H Lee
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.18363v1 Announce Type: new Abstract: Feed-forward networks hold two thirds of a transformer's non-embedding parameters, yet the architecture has not received a necessity test that controls parameters, compute, and depth at once. We pretrain attention-only decoder transformers (Simple Atte...
35. Scalable and Efficient Joint Spiking Embedding Predictive Architecture for Large-Scale Dynamic Graphs ​
Author: Huizhe Zhang, Yuchang Zhu, Huazhen Zhong, Liang Chen, Zibin Zheng
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2607.18412v1 Announce Type: new Abstract: Dynamic graph learning aims to capture evolving structural and semantic patterns in real-world systems, such as fraud detection and recommender systems. Due to the scarcity of labeled data in real-world dynamic graphs, recent studies have introduced ge...
36. PAC--Bayes Bounds on Quotient Parameter Spaces: Geometry-induced Implicit-Bias Priors ​
Author: Nicola Aladrah, Fabio Anselmi
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.18422v1 Announce Type: new Abstract: Overparameterized models often have continuous parameter symmetries, so different parameters define the same predictor. We show that PAC--Bayesian analysis should be performed on the quotient predictor space: pushing a prior and posterior to the quotie...
37. Intelligence from Learnable Novelty ​
Author: Yanbo Zhang, Michael Levin
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, nlin.AO
arXiv:2607.18433v1 Announce Type: new Abstract: Intelligence appears under different names in different fields: as data compression in statistics and machine learning, as universal computation in dynamical systems, and as adaptive behavior in agents. Each field carries its own objective, and the two...
38. CANDOR: Chance-Calibrated Discordance in Frozen Foundation Encoders ​
Author: Soroosh Tayebi Arasteh, Sven Nebelung, Daniel Truhn
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CV
arXiv:2607.18451v1 Announce Type: new Abstract: Frozen encoders are chosen by how well a lightweight head reads a finding from their features, not whether the geometry separates it. Nearest-neighbor discordance does, but with unequal banks the opposite-label neighbor wins on density, not geometry, s...
39. Estimating Rare Events in Language Models with Proper Evaluation ​
Author: Nikita Y. Parulekar, Anqi Liu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18454v1 Announce Type: new Abstract: Quantifying the risk of rare failures in language models, such as those triggered by adversarial distribution shifts or very large-scale deployments, requires estimating probabilities far too small for random sampling. While recent work has formalized ...
40. AHEAD: Advancing Multi-Class Label Aggregation with Interpretable Cross-Annotator Modeling ​
Author: Ju Chen, Sijia Xu, Jun Feng, Zhiqiang Gao, Zhengyi Yang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18465v1 Announce Type: new Abstract: Crowdsourced labeling provides valuable labeled data for domains across natural language processing, computer vision, and video. Label aggregation aims to infer latent true labels from noisy and biased annotations, with the key lying in annotator relia...
41. Weak-to-Strong Learning in Decision Making ​
Author: Jingwei Ji, Renyuan Xu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18467v1 Announce Type: new Abstract: Many operational decisions rely on predictive models that estimate uncertain outcomes conditional on observable contexts. Training such models, however, often faces a fundamental data asymmetry: labeled outcomes are scarce or costly to obtain, while co...
42. RRPO: Reference-Relative Policy Optimization with Stratified Conditional Rollouts ​
Author: Yuxin Xiong, Xunyi Jiang, Rohan Surana, Xintong Li, Sheldon Yu, Nikki Lijing Kuang, Ryan A. Rossi, Jingbo Shang, Tong Yu, Julian McAuley, Junda Wu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18470v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has shown strong effectiveness in reinforcement learning from verifiable feedback, where sampled rollouts can be compared within a group using task-provided correctness signals. However, extending group-relativ...
43. Hybrid Latent-Structural Fusion (HLSF) for Cyber Anomaly Detection ​
Author: Dorianis M. Perez, Maksim E. Eren, Bryan E. Kaiser
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18479v1 Announce Type: new Abstract: Malicious anomalous activity detection is a fundamental challenge for cyber security systems. Both tensor decomposition under statistical framework with CANDECOMP-PARAFAC alternating Poisson regression (CP-APR) and normalizing flows have proven to be p...
44. Attractor Geometry Determines the Identifiability Limits of System Discovery ​
Author: Matteo Gallo, Fabio Anselmi, Paolo Lazzari
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, nlin.CD
arXiv:2607.18490v1 Announce Type: new Abstract: Symbolic discovery of governing equations from data is limited not only by algorithm design and data volume, but by the geometry of the attractor: what the long-run dynamics allow to be recovered. Using a within-system design on Lorenz-84, where one fo...
45. Now We Know? A Systematic Comparison of TerraMind and THOR ​
Author: Frederick Schindlegger, Kenzo Bounegta, Eva Gmelich Meijling, Johannes Jakubik, Arnt-B{\o}rre Salberg, Theodor Forgaard, Nicolas Longepe, Valerio Marsocci
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2607.18504v1 Announce Type: new Abstract: Benchmarks for Geospatial Foundation Models (GFMs) increasingly rank models by aggregate score, but such rankings obscure why models differ: how much of the gap is architecture, how much is decoder capacity, and how much is a use-case-specific artefact...
46. Automated Data Engineering and Feature Selection for the Case Study of Warpage Detection in Fused Deposition Modeling ​
Author: Saleh Valizadeh Sotubadi, Nazanin Mahjourian, Vinh Nguyen
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18515v1 Announce Type: new Abstract: This study contributes toward development of an Automated Data Processing (ADP) framework designed to evaluate and reinforce optimal machine learning model-feature combinations for predictive tasks in fused deposition modeling (FDM) process datasets. T...
47. Signed Rectified Flow: Negativity-Controlled Generation ​
Author: Runlong Liao, Baiyu Su, Lizhang Chen, Qiang Liu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.18516v1 Announce Type: new Abstract: We introduce Signed Rectified Flow (Signed RF), a generalization of Rectified Flow that targets the signed measure $\pi^{sign} = (1+\alpha)\pi^+ - \alpha\pi^-$, where $\alpha>0$, $\pi^+$ is the distribution to promote, and $\pi^-$ is the distribution t...
48. Adaptive Two-Stage Online Learning for Service-Affecting Failure Detection in Mobile Core Networks ​
Author: J. du Toit, G. Fita, J. Salzwedel, A. Stoltz, R. Wolhuter
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18522v1 Announce Type: new Abstract: Mobile network operators monitor aggregated traffic volumes to assess the operational health of core network infrastructure. Reliable failure detection is challenging due to strong temporal structure, non-stationarity, measurement artefacts, and extrem...
49. Censoring-Aware In-Context Learning for Generalized Supplier Lead Time Estimation in Supply Chain Planning ​
Author: Christopher Wang, Sebastien Ouellet, Behrouz Haji Soleimani, Ali Etemad
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18530v1 Announce Type: new Abstract: Supplier lead time forecasting is a central input to material requirements planning, inventory optimization, and supply chain risk management. However, many industrial lead time datasets are naturally right-censored: at the time forecasts are required,...
50. Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary ​
Author: Jan Kirin
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18553v1 Announce Type: new Abstract: Can a language model read the quality of ongoing computation, and can an external intervention turn that readout into better outcomes? We test both questions in a frozen 2.6B looped transformer, Ouro-RLTT. On GSM8K, a strict pre-answer probe excludes t...
51. Robust Multi-View Classification under Noisy Supervision via Global Anchor Consensus ​
Author: Yuliang Yang, Hongzhe Zhang, Huiru Wang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.18561v1 Announce Type: new Abstract: In recent years, multi-view learning has attracted increasing attention, as it integrates the complementary information of heterogeneous views. Most existing multi-view classification methods rely on accurate annotations to guarantee performance. Howev...
52. AMICA-Python: Adaptive Mixture Independent Component Analysis with Anderson Acceleration ​
Author: Scott Huberty, Christian O'Reilly
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18568v1 Announce Type: new Abstract: Adaptive Mixture Independent Component Analysis (AMICA) is widely used in EEG research and has long been associated with strong empirical performance for blind source separation. Despite its impact, practical use has historically depended on a single F...
53. Conditioned Direct Feedback Alignment via Activity and Error Geometry ​
Author: Houman Safaai, Varun Reddy, Bernardo L. Sabatini
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, q-bio.NC
arXiv:2607.18574v1 Announce Type: new Abstract: Direct feedback alignment (DFA) trains hidden layers with fixed random projections of the output error, avoiding the transposed-weight backward pass of backpropagation (BP). We study a failure mode of DFA training that is distinct from feedback quality...
54. On the Diverse Dynamical Behaviors Arising in Deep Linear Transformers ​
Author: Sixu Li, Thomas Jacob Maranzatto, Jan Peszek, Trevor Teolis, Semih Akkoc, Konstantin Riedl, Sennur Ulukus, Nicol'as Garc'ia Trillos
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, math.DS
arXiv:2607.18584v1 Announce Type: new Abstract: We study the inference-time behavior of deep linear encoder-only transformers through the lens of interacting particle systems. In this perspective, tokens are modeled as particles that interact dynamically through successive linear self-attention laye...
55. Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States ​
Author: Armin Sommer
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18589v1 Announce Type: new Abstract: Reinforcement learning is conventionally divided into model-based and model-free methods. In this taxonomy, model-based methods perform lookahead planning over a learned world model, whereas model-free methods learn a reactive state-action mapping. Rec...
56. A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space ​
Author: Shuangyao Huang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.MA
arXiv:2607.18597v1 Announce Type: new Abstract: Counterfactual credit assignment has proven effective in multi-agent reinforcement learning (MARL) for discrete action spaces, yet its extension to continuous-action cooperative tasks remains challenging. Existing methods that approximate the counterfa...
57. BRIDGE: Bottleneck-Aware Regulator-Set Inference and Diagnosis for Cooperative Gene Regulatory Recovery ​
Author: Maryam Rahimimovassagh, Clayton Thomas Barham, Ivan Garibay, Niloofar Yousefi
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18602v1 Announce Type: new Abstract: Cooperative gene regulation often depends on groups of regulators acting jointly, but most gene regulatory network (GRN) inference methods output pairwise regulator-target rankings. We introduce Bottleneck-Aware Regulator-Set Inference and Diagnosis (B...
58. Graph Neural Network-based Algorithm Selection for the Traveling Salesman Problem: A Systematic Study of Cost and Rank Losses under Distinct Budget Regimes ​
Author: Zhaoxuan Li, Jiale Yang, Yifei Lu, Mustafa Misir
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18632v1 Announce Type: new Abstract: Automated Algorithm Selection (AS) aims to improve problem-solving performance by selecting, for each problem instance, the most suitable algorithm from a predefined portfolio. This is particularly relevant to the Traveling Salesman Problem (TSP), wher...
59. Mark, Don't Erase: Token Inoculation for Dual-Use Knowledge in LLMs ​
Author: Seunghyun Lee, Dongyoon Han, Sangdoo Yun
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.18639v1 Announce Type: new Abstract: Safety interventions on dual-use knowledge typically choose between destroying hazardous content (e.g., unlearning, filtering) and suppressing it at the output layer (e.g., refusal training); both pay a tax in adjacent-domain competence or over-refusal...
60. Spaghetti Architect: A Contamination-Resistant, By-Construction-Labelled, Multi-Language Code Dataset Generator ​
Author: Yuxiang Ji
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.SE
arXiv:2607.18642v1 Announce Type: new Abstract: Mined code corpora are abundant but uncontrolled: a snippet's semantics, surface "messiness," and difficulty are whatever the wild contained; there is no known-optimal reference to grade against; and any public sample may already sit in a model's train...
61. Exposure-Based Reinforcement Learning to Rank ​
Author: Harrie Oosterhuis, Rolf Jagerman, Zhen Qin, Xuanhui Wang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2607.18689v1 Announce Type: new Abstract: Reinforcement learning (RL) methods for learning-to-rank (LTR) can optimize (almost) any ranking goal, e.g., from precision or discounted cumulative gain to fairness-of-exposure or ranking distillation. However, standard RL is ineffective and computati...
62. Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning ​
Author: Junyao Yang, Yucheng Shi, Zongxia Li, Zhongzhi Li, Ruhan Wang, Xiangxin Zhou, Kishan Panaganti, Haitao Mi, Leowei Liang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.18722v2 Announce Type: new Abstract: Asynchronous reinforcement learning improves throughput by decoupling rollout generation from optimization, but staleness is an inevitable byproduct compounded by policy lag, engine delays, and mixture-of-experts routing. From a trust-region perspectiv...
63. Contraction-Gauge Preconditioning for Quantized Matrix Multiplication ​
Author: Piyush Sao, Narasinga Miniskar, Pedro Valero-Lara, Keita Teranishi, Sudip Seal
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, cs.NA, math.IT, math.NA
arXiv:2607.18745v1 Announce Type: new Abstract: We study low-precision computation of C=AB with both factors quantized. We derive an exact finite-dimensional identity for the expected squared product error under independent, zero-mean entrywise errors with known variance fields; it holds exactly for...
64. ConceptCF: Concept-based Counterfactuals for the Explainability of Time Series ​
Author: Annemarie Jutte, Faizan Ahmed, Jeroen Linssen, Maurice van Keulen
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18748v1 Announce Type: new Abstract: This paper proposes ConceptCF, a method for counterfactual generation that operates on human-interpretable concepts. In high-stakes domains such as healthcare and predictive maintenance, artificial intelligence models can increase efficiency and safety...
65. Is EEG-to-Text Feasible in Real-World Scenarios? An In-Depth Analysis Using a Neuropsychology-Inspired Benchmark ​
Author: Zihan Zhang (Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology), Yu Bao (Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology, Shanghai Innovation Institute), Xiao Ding (Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology), Tianyi Jiang (State Key Laboratory for Novel Software Technology, Nanjing University), Kai Xiong (Zhongguancun Laboratory)
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.ET, q-bio.NC
arXiv:2607.18749v1 Announce Type: new Abstract: Translating brain signals into text could restore communication for people with severe paralysis, yet practically usable systems to date rely on invasive electrocorticography (ECoG). Electroencephalography (EEG) offers a non-invasive alternative, and E...
66. Decafs: Disentangled Conditional adversarial Flows ​
Author: Anirudh jain, Sakshi Varshney, Samuel Kaski, Vikas Garg
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18755v1 Announce Type: new Abstract: Flow-based models have established state-of-the-art performance in generative modeling across domains, but are hard to interpret due to their complex latent embeddings. In particular, the entanglement of generative factors in the latent space hinders c...
67. Relative Positions Generalize, Absolute Positions Memorize: An Implicit-Bias Account of Length Generalization in Attention ​
Author: Subham Singh, Ashutosh Mishra, Subha Raut
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18759v1 Announce Type: new Abstract: Transformers with relative positional encodings often extrapolate to sequences longer than those seen during training, whereas transformers with learned absolute encodings typically do not. This is a robust empirical regularity, and the explanations of...
68. Formulation-Level Auto-Tuning for QUBO-Based Machine Learning: A Case Study Across Multiple Quantum-Inspired Annealers ​
Author: Naoya Mizuki, Takahiro Katagiri, Daichi Mukunoki, Tetsuya Hoshino
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.PF
arXiv:2607.18774v1 Announce Type: new Abstract: This paper presents an Optuna-based formulation-level auto-tuning framework for support vector machines (SVMs) implemented on multiple quantum-inspired annealers. In an annealing-based SVM, continuous dual variables are discretized and converted into a...
69. PertReason: A Knowledge-Grounded Benchmark and Framework for Cell-State-Conditioned Mechanistic Reasoning of Perturbation Effects ​
Author: Dongkwan Kim, Yiming Gao, Yining Yang, Yang Shen
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, q-bio.MN
arXiv:2607.18777v1 Announce Type: new Abstract: Evaluating machine learning in scientific domains requires separating correct predictions from correct reasons under realistic distribution shifts. We introduce PertReason, a knowledge-grounded benchmark and framework suite for cell-state--conditioned ...
70. QScheduler: Adaptive Gradient Sampling for Zeroth-Order On-Device Training on INT8 NPUs ​
Author: Victor Felipe Domingues Do Amaral (GeePs), Pierre Demaj (GeePs), Erwan Libessart (GeePs), Laurent Folliot (GeePs), Anthony Kolar (GeePs), Philippe B'enab`es (GeePs)
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18802v1 Announce Type: new Abstract: Zeroth-Order (ZO) optimization enables On-Device Learning (ODL) on NPU-equipped microcontrollers by estimating gradients through forward passes alone, bypassing the need for backpropagation primitives and reducing memory requirements. The number of gra...
71. Elicitation without Backpropagation: Steering Model Behavior by Optimizing the Latent Posterior ​
Author: Garrett Baker, Vinayak Pathak, Daniel Murfet, Susan Wei
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.18804v1 Announce Type: new Abstract: In the \emph{latent posterior model} of transformer behavior, the next-token distribution arises from a posterior over latent predictive models conditioned on the context, mixed to generate continuations. We exploit this model in settings where it is e...
72. Countercurrent Multiplier Networks: A Renal-Inspired Iterative Operator with Provably Bounded Fixed-Point Dynamics ​
Author: Snigdha Chandan Khilar
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP
arXiv:2607.18829v1 Announce Type: new Abstract: The mammalian kidney concentrates urine using a mechanism with no analogue in current neural architectures: the countercurrent multiplier. Two anti-parallel flows joined at a hairpin recirculate a weak magnitude-bounded local pump into a large axial gr...
73. From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning ​
Author: Garvit Singla, Uma Maheswari Natarajan, Raghuram Bharadwaj Diddigi
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18830v1 Announce Type: new Abstract: Model-Agnostic Meta-Learning (MAML) is a widely used framework for reinforcement learning (RL) that enables efficient transfer by learning global policy parameters that can be rapidly adapted to new tasks. MAML training proceeds in two loops: an inner ...
74. ABOPD: Antibody CDR Design via On-Policy Distillation ​
Author: Zhuo Yang, Jiaying He, Jiaqing Xie, Daolang Wang, Xipeng Qiu, Yuxin Wang, Tianfan Fu, Beilun Wang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18835v1 Announce Type: new Abstract: Antibodies are essential therapeutic molecules, and their complementarity-determining regions (CDRs) form the primary antigen-recognition interface. Recent protein generative models have demonstrated broad capabilities in biomolecular design, yet post-...
75. Regime-Aware Physics-Guided Early Warning of Lithium-Ion Battery Thermal Runaway Using Thermo-Mechanical Signals ​
Author: Syed Sajid Ullah, Muhammad Zunair Zamir, Salman Khan
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18860v1 Announce Type: new Abstract: Thermal runaway in lithium-ion batteries poses a major safety risk to electric vehicles and energy storage systems. Current early-warning methods depend mainly on temperature and may therefore miss mechanical precursors that emerge before rapid heating...
76. HindsightBench: A Black-Box Behavioral Audit Protocol for Parametric Hindsight in Time-Indexed LLM Decision Tasks ​
Author: Haozhe Jia
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.18867v1 Announce Type: new Abstract: Large language models leak parametric knowledge of realized outcomes into historical financial decision tasks. Existence is settled; what users lack is a cheap way to audit a given model for it. We present HindsightBench, a black-box behavioral audit p...
77. Reinforcement Learning for Delivery Drone-Based Participatory Sensing in Dynamic Environments ​
Author: Xin Ouyang, Songxin Lei, Xusen Guo, Yutian Jiang, Sijie Ruan, Yuxuan Liang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CY
arXiv:2607.18874v1 Announce Type: new Abstract: Using Unmanned Aerial Vehicle (UAV) for urban sensing has emerged as a powerful paradigm to monitor the status of the city, e.g., air quality and noise levels, through agile aerial crowdsourcing. Despite this potential, existing UAV-based sensing appro...
78. Physics-Informed Super-Resolution of Atmospheric Data ​
Author: Chang Xu, Gencer Sumbul, Hugo Porta, Manon B'echaz, Sebastian Schemm, Devis Tuia
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18877v1 Announce Type: new Abstract: In the context of global warming, extreme events have become more frequent and intense, making their trustworthy detection and forecasting more important than ever. Yet, atmospheric observations lack sufficient spatial resolution, motivating atmospheri...
79. RAMP: Recognition parametrisation by Amortised Message Passing ​
Author: Lior Fox, Kai Biegun, James Heald, Samo Hromadka, Arielle Rosinski, Maneesh Sahani
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18883v1 Announce Type: new Abstract: A central aim of unsupervised learning is to uncover latent factors that explain dependencies among observations. Probabilistic models typically achieve this by introducing multiple latent variables linked through a graph of conditional relationships, ...
80. KALE: Kernel Alignment with Loss Equilibration for Stable CLIP-DINOv2 Alignment at Web Scale ​
Author: Micha{\l} Paw{\l}owicz
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18885v1 Announce Type: new Abstract: Kernel-based alignment of CLIP toward a vision centric teacher such as DINOv2 (KUEA) improves CLIP's visual representations while preserving text-encoder compatibility, using a fixed trade-off weight tuned on curated ImageNet-1K. We ask whether this tr...
81. Breaking Feedback-Blindness: Utility-Augmented Transformer for Sequential Decision Making ​
Author: Yuyang Shen, Shan Dai, Daimin Chen
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18910v1 Announce Type: new Abstract: Sequential decision making in non-stationary and partially observable environments requires rapid adaptation to latent regime changes. However, existing Transformer decision models face a structural bottleneck in the retrieval mechanism: even when rewa...
82. Circuit Claims Depend on What Is Extracted and How It Is Compared ​
Author: Yang Sheng, Jie Fu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18921v1 Announce Type: new Abstract: Circuit extraction identifies a small set of model components whose presence preserves a target behavior under ablation, and the resulting circuit is often read as the mechanism behind that behavior. We argue that this reading is under-determined: pres...
83. Visual Semantic Decoding of Electrocorticography from Video Stimuli using End-to-End Deep Learning ​
Author: Stella Ho, Joel Villalobos, Joseph West, Jingyang Liu, Weijie Qi, Haruhiko Kishima, Ryohei Fukuma, Takufumi Yanagisawa, Sam E. John, David B. Grayden
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC
arXiv:2607.18923v1 Announce Type: new Abstract: ECoG-based visual semantic decoding enables inference of semantic interpretation of visual perception from complex, noisy brain activity. This study examines the feasibility of visual semantic decoding using an end-to-end deep learning framework using ...
84. Functional Equivalence and Geometric Diversity in Neural Network Approximations: An Empirical Characterization ​
Author: Anuragine S A, Prem Jagadeesan
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18930v1 Announce Type: new Abstract: The Universal Approximation Theorem states that a neural network with a single hidden layer is sufficient to approximate any continuous univariate function on a compact domain to arbitrary error. However, the uniqueness of such neural network represent...
85. H$^2$SD: Hybrid Hindsight Self-Distillation ​
Author: Qiye Cai, Yichuan Ma, Linyang Li, Peiji Li, Yongkang Chen, Qipeng Guo, Yicheng Zou, Xiaocheng Feng, Bing Qin
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.18955v2 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) provides reliable outcome supervision for language model reasoning, but a scalar trajectory reward offers limited token-level guidance. Existing self-distillation methods add a privileged teacher bu...
86. SFGA: A Statistics-First Gating Architecture with Adjudicative Escalation for Trustworthy SFT Data Procurement ​
Author: Arther Tian, Alex Ding, Simon Wu, Aaron Chan
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2607.18960v1 Announce Type: new Abstract: Procuring supervised fine-tuning (SFT) data forces a buyer to decide, before any downstream training, whether a candidate corpus is worth acquiring. We present \sys{}, a statistics-first gating architecture that treats procurement as a cost-aware routi...
87. Variational meta-learning inference for low dimensional neural system identification ​
Author: Matteo Rufolo, Dario Piga, Marco Forgione
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY
arXiv:2607.18965v1 Announce Type: new Abstract: Deep learning has proven highly effective for nonlinear system identification, but heavily parameterized neural networks are prone to overfitting in low-data regimes and lack reliable uncertainty quantification. The recently developed manifold meta-lea...
88. Subject-Conditioned Glucose Forecasting in Type-1 Diabetes ​
Author: Giorgia Rigamonti, Mirko Paolo Barbato, Davide Marelli, Paolo Napoletano
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2607.19006v1 Announce Type: new Abstract: Accurate forecasting of blood glucose concentration is key in the management of Type 1 Diabetes, facilitating early detection of adverse glycemic events and supporting timely therapeutic interventions. Despite recent advances in glucose prediction, mos...
89. Biological Amnesia in ICU Time-Series Prediction: A Drift-Adaptive Two-Stream Architecture with Temporal Retrieval ​
Author: Fatema Ferdous Tamanna, K. M. Merajul Arefin, Md. Abdul Masud
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR, q-bio.QM
arXiv:2607.19020v1 Announce Type: new Abstract: Background: Clinical decision support systems degrade silently as treatment protocols evolve, yet standard adaptation methods treat models as monolithic blocks, unable to distinguish stable patient physiology from shifting institutional practice. Metho...
90. Unsupervised Multi-kernel Learning for Automated Algorithm Selection ​
Author: Yihang Lu, Tome Eftimov, Carola Doerr
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.19031v1 Announce Type: new Abstract: Automated algorithm selection in black-box optimization typically relies on supervised models that map landscape features to algorithm performance labels. Such models are costly to train, benchmark-dependent, and often fail to generalize to unseen prob...
91. Spectral Higher-Order Neural Networks Have Sharp Expressivity Bounds ​
Author: Gianluca Peri, Diego Febbe, Duccio Fanelli
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19042v1 Announce Type: new Abstract: Neural hypergraphs are a natural generalization of neural networks, the reference models in modern machine learning. Yet, their deployment has proven demanding: the number of weighted hyperedges required leads to an intractable parameter explosion. How...
92. Adopting Reinforcement Learning with Verifiable Rewards for Molecular Generation ​
Author: Mingxuan Ouyang, Hao Lan, Wanyu Lin
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.19044v1 Announce Type: new Abstract: Leveraging large language models (LLMs) for molecular generation has shown remarkable potential in chemical and drug design. Current methods primarily rely on supervised training or fine-tuning with limited datasets, which are insufficient to capture c...
93. Probabilistic Physics-Aware Machine Learning Predictions of Electric Truck Energy Consumption with Field Data ​
Author: Hannes Nilsson, Rafael Basso, Bal'azs Kulcs'ar, Morteza Haghir Chehreghani
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.19054v1 Announce Type: new Abstract: In this work, we incorporate first principle physics into the construction of data-driven methods by considering a model that accounts for the different sources of energy losses during vehicle operations. Our results show that Bayesian linear regressio...
94. Where Should Optimizer State Live? Tiered State Allocation for Memory-Efficient Mixture-of-Experts Training ​
Author: Nuemaan Malik
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19058v1 Announce Type: new Abstract: Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps 50.6 GB of first and second moments to update 12.6 GB of bfloat16 weights. We study SkewAdam...
95. Deep learning-based prediction of time-resolved adhesive forces in viscoelastic Hertzian contacts ​
Author: Ali Maghami, Merten Stender, Michele Ciavarella, Antonio Papangelo
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.soft, cs.AI, physics.data-an, stat.ML
arXiv:2607.19060v1 Announce Type: new Abstract: Fast prediction of the response of adhesive soft viscoelastic contacts represents a current challenge in soft robotics and for gripping and manipulation tasks. Determining the complete time-resolved force trajectory requires full numerical simulations,...
96. GEqTrain: A Configuration-Driven Framework for Retargeting Equivariant Graph Neural Networks Across 3D Scientific Tasks ​
Author: Daniele Angioletti, Marco Nobile, Vittorio Limongelli
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM
arXiv:2607.19083v1 Announce Type: new Abstract: Equivariant graph neural networks provide a powerful modeling language for three-dimensional scientific data, but their reuse is often limited by implementations tied to specific tasks, outputs, and training regimes. We present GEqTrain, a configuratio...
97. An unsupervised clustering analysis of breast cancer data derived from electronic health records enhanced through UMAP dimensionality reduction ​
Author: Davide Chicco, Nicoletta Benvenuto
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2607.19089v1 Announce Type: new Abstract: Breast cancer is one of the most widespread types of cancer, affecting approximately 8 million women worldwide. Electronic health records of patients diagnosed with this disease can serve as valuable datasets for computational analyses, enabling the di...
98. Predicting Activities in Aqueous Electrolyte Solutions with Hybrid Machine Learning ​
Author: Zeno Romero, Maximilian Kohns, Fabian Jirasek
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph
arXiv:2607.19114v1 Announce Type: new Abstract: Activities in aqueous electrolyte solutions, usually described by ionic activity and osmotic coefficients, are important properties for modeling many processes in industry and nature. Established activity models, such as those of Pitzer or Bromley, req...
99. Parallel Noising in Neural Markov Logic Networks ​
Author: Peter Jung, Giuseppe Marra, Ondrej Kuzelka
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19126v1 Announce Type: new Abstract: Neural Markov Logic Networks (NMLNs) are a flexible neurosymbolic relational model. Previous work has shown that, although NMLNs achieve strong performance as generative models for small relational structures, they underperform diffusion-based generati...
100. One Model, Many Graphs: Learning over Attributed Graphs across Heterogeneous Modalities with Vision-Language Models ​
Author: Jiayi Yang, Yifang Chen, Yuanfu Sun, Jiajin Liu, Qiaoyu Tan
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.19128v1 Announce Type: new Abstract: Vision-language models (VLMs) provide a unified representation space for textual and visual information, yet their potential as general-purpose backbones for graph-structured data remains largely unexplored. In practice, attributed graphs exhibit subst...
101. Incomplete Observations Boost Evolutionary Performance in Ocean Modeling ​
Author: Yangyang Kong, Yutong Jiang, Yanhai Gan, Junyu Dong, Feng Gao, Xiaopei Lin
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19147v1 Announce Type: new Abstract: Data-driven methods have revolutionized ocean modeling, yet current approaches rely heavily on complete reanalysis datasets, imposing computational constraints and limiting model performance to that of the training data. Here, we present a generative s...
102. Breaking the Homogeneity Assumption: Specialized Multi-Generator Adversarial Learning for Rare Failure Detection in Predictive Maintenance ​
Author: Alexis Lazanas, Georgios Kampouropoulos
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19153v1 Announce Type: new Abstract: Supervised learning models in the predictive maintenance field are regularly trained on highly imbalanced industrial datasets: machine failures occur rarely but have a disproportionate effect on operations. In addition to the clear class disparity, fai...
103. Neural Kolmogorov Equations: Parallelizable Learning of Stochastic Dynamics under General Noise ​
Author: Arthur Bizzi, Olga Fink
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.19173v1 Announce Type: new Abstract: Neural stochastic differential equations (SDEs) have emerged as powerful tools for learning noisy or stochastic dynamics directly from data; however, existing approaches largely assume uncoupled and continuous noise, limiting their applicability to rea...
104. Conservative Query and Adaptive Regularization for Offline RL Under Uncertainty Estimation ​
Author: Li-Rong Zhou, Qin-Wen Luo, Sheng-Jun Huang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.19199v1 Announce Type: new Abstract: Offline reinforcement learning (RL) aims to learn an effective policy from a static dataset, but its performance is fundamentally limited by dataset coverage. Action preference queries leverage expert feedback without additional environment interaction...
105. AdaFlash: Adaptive Speculative Decoding via On-Policy Distilled Diffusion Drafters ​
Author: Yu-Yang Qian, Hao-Cong Wu, Chen Chen, Jiacheng Sun, Zhenhua Dong, Peng Zhao, Zhi-Hua Zhou
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.19223v1 Announce Type: new Abstract: Speculative decoding, in which a lightweight draft model first generates a draft sequence that is then verified in parallel by the target model, has become a prevalent paradigm for accelerating large language model inference. Recent work such as DFlash...
106. S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning ​
Author: Kshitij Kumar Srivastava, Kshitij Jerath
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.MA
arXiv:2607.19232v1 Announce Type: new Abstract: Hierarchical Reinforcement Learning (HRL) intends to separate strategic planning from primitive execution. It has been widely successful in solving long-horizon and complex tasks, where flat-RL algorithms have difficulty in learning. However, while the...
107. In-Context Time Series Classification with Random Convolutional Features ​
Author: Joscha C"uppers, Jilles Vreeken
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.19234v1 Announce Type: new Abstract: Time series classification is central to domains like medical signal analysis, industrial monitoring, and sensor-based activity recognition, where class information manifests as localized shapes, specific frequencies, temporal shifts, or complex cross-...
108. DBMol: Design of High-Affinity, Target-Specific Small Molecules through Structure Prediction Models ​
Author: Yiming Qin, Kai Yi, Miruna Cretu, Sjors H. W. Scheres, Pietro Li`o, Pascal Frossard
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.19237v1 Announce Type: new Abstract: Designing small molecule ligands that bind with high affinity to specific protein pockets is a fundamental goal in drug discovery, as small molecules constitute a major fraction of approved therapeutics. Recent breakthroughs in structure prediction, su...
109. Thermodynamics-Informed Input Reparameterization for Neural Prediction of Real-Fluid Thermodynamic Properties in Supercritical Combustion ​
Author: Haoze Zhang, Han Li, Ke Xiao, Yangchen Xu, Runze Mao, Zhi X. Chen
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph, physics.flu-dyn
arXiv:2607.19241v1 Announce Type: new Abstract: Real-fluid thermodynamic property evaluation is a major computational cost in supercritical combustion simulations. In the enthalpy-based pressure-correction formulation, the closure evaluates temperature T, density $\rho$, and compressibility coeffici...
110. Benchmarking Generalization in Financial Statement Fraud Detection: robust evaluation and novel tasks ​
Author: Guy Stephane Waffo Dzuyo (Forvis Mazars, LORIA CNRS Universit'e de Lorraine), Ga"el Guibon (LORIA CNRS Universit'e de Lorraine, LIPN CNRS Universit'e Sorbonne Paris Nord), Christophe Cerisara (LORIA CNRS Universit'e de Lorraine), Luis Belmar-Letelier (Forvis Mazars)
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19259v1 Announce Type: new Abstract: Financial statement fraud detection (FSFD) is crucial for market integrity but faces challenges from increasingly sophisticated schemes and under-utilized textual data in financial reports. Existing methods often rely on random data splits, leading to ...
111. Toward Auditable Fraud Detection: Combining Graph Features, Model Explanations, and Agentic Case Investigation ​
Author: Rahil Sharma
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19266v1 Announce Type: new Abstract: Fraud detection systems must scale with rising transaction volume while remaining explainable and reviewable. We study a layered pipeline on the PaySim dataset that combines a gradient-boosted classifier, graph-derived structural features, an autoencod...
112. GUIDED Network-Agnostic Feature Initialization for Spatial Transferability in GNN-based Models ​
Author: Alessandro Scalese, Santhanakrishnan Narayanan, Constantinos Antoniou
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19270v1 Announce Type: new Abstract: The Traffic Assignment Problem is a fundamental but computationally expensive component of transportation planning. While Graph Neural Networks have emerged as fast, data-driven surrogates, their practical deployment is severely constrained by a spatia...
113. A Reinforcement-Learning-Augmented Liquid-Fueled Reactor Network Model for Predicting Lean Blowout in Gas Turbine Combustors ​
Author: Philip John, Eloghosa Ikponmwoba, Pinaki Pal, Opeoluwa Owoyele
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.19281v1 Announce Type: new Abstract: This study introduces a reinforcement learning (RL) framework for generating optimal liquid-fueled reactors to improve lean blowout (LBO) predictions in gas turbine combustors. Existing approaches for determining cluster boundaries rely on manual heuri...
114. Real-time optimal control with shallow recurrent decoder networks ​
Author: Matteo Tomasetto, Francesco Braghin, J. Nathan Kutz, Andrea Manzoni
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2607.19302v1 Announce Type: new Abstract: Controlling dynamical systems in real-time across multiple scenarios is critical to enabling adaptive control strategies, ensuring stability and efficiency. However, to tailor control actions in response to varying scenarios, traditional optimal contro...
115. Riemannian Deep Learning:Modules, Networks, and Geometries ​
Author: Chen Ziheng
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DG
arXiv:2607.19305v1 Announce Type: new Abstract: Deep neural networks on manifold-valued representations have attracted growing interest, but many basic components remain tied to specific manifolds, rely on Euclidean approximations, or require costly and numerically fragile geometric operations. This...
116. Staypoint Detection from Noisy Trajectory Data [Experiment Paper] ​
Author: Lance Kennedy, Hossein Amiri, Yueyang Liu, Riyang Bao, Hanqi Chen, Mohammad Hashemi, Ruochen Kong, Xiaotong Liu, Joon-Seok Kim, Shengpu Tang, Liang Zhao, Andreas Z"ufle
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CG
arXiv:2607.19312v1 Announce Type: new Abstract: Detecting staypoints from raw trajectory data is fundamental to numerous spatial computing applications. This process transforms raw numeric sequences of geolocations into semantically meaningful locations, such as homes, workplaces, or restaurants. De...
117. Off-Context GRPO: Learning to Reason on Hard Problems using Privileged Information ​
Author: Priyank Agrawal, Ankur Samanta, Shervin Ghasemlou, Jalaj Bhandari, Kavosh Asadi, Daniel Jiang, Aditya Modi
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19313v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves reasoning in large language models. Yet, typical RLVR approaches fail on difficult problems: when a model cannot generate any correct solutions, it receives \textit{zero} learning signal. P...
118. CircuitKIT : Circuit Discovery, Evaluation, and Application Toolkit for Mechanistic Interpretability ​
Author: Pratinav Seth, Hem Gosalia, Aditya Kasliwal, Vinay Kumar Sankarapu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.ET
arXiv:2607.19317v1 Announce Type: new Abstract: Circuit analysis can support not only model explanation but also downstream interventions such as pruning, editing, steering, and selective fine-tuning. However, conducting such analyses currently requires stitching together separate implementations fo...
119. ISO: An RLVR-Native Optimization Stack ​
Author: Hanqing Zhu, Wenyan Cong, Zhizhou Sha, Sagnik Mukherjee, Xinyuan Song, David Gonz'alez-Mart'inez, Xiaoxia Wu, Yuandong Tian, Shiwei Liu, David Z. Pan, Zhangyang "Atlas" Wang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19331v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is rapidly advancing the reasoning capabilities of language models, yet the optimization layer that converts reward feedback into weight-space updates remains poorly understood. Building on our prio...
120. ROMS-IMLE: A Minimalist Approach to Competitive Single-Step Generative Modelling ​
Author: Chirag Vashist, Ke Li
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.19332v1 Announce Type: new Abstract: Generative models have undergone many generations of evolution, from VAEs/GANs to diffusion/flow matching. Along the way, the underlying techniques have become more complicated and various beliefs about what drives strong empirical performance have tak...
121. Provable diffusion-based posterior sampling for linear inverse problems via DDIM ​
Author: Yuchen Jiao, Na Li, Changxiao Cai, Yuxin Chen, Gen Li
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.19333v1 Announce Type: new Abstract: Diffusion-based methods have achieved remarkable empirical success in solving inverse problems. However, many existing posterior samplers either lack rigorous theoretical guarantees or incur substantial computational overhead. We propose a simple and e...
122. Integro-differential equations in angular stabilization of drone motion by distributed feedback control ​
Author: Alexander Domoshnitsky, Oleg Kupervasser, Anatoly Polonsky
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, math.OC
arXiv:2607.18251v1 Announce Type: cross Abstract: In this paper, we propose angular stabilization of drone motion using distributed feedback control in the form of an integral operator. It should be stressed that the memory of this integral operator could be unbounded. It is intuitively clear that l...
123. Cross-Dialect Generalization Without Retraining: Benchmarks and Evaluation of Schema-Derived Constrained Decoding for MLIR ​
Author: Plawan Kumar Rath
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.PL
arXiv:2607.18254v1 Announce Type: cross Abstract: Multi-Level Intermediate Representation (MLIR) underlies modern ML compiler infrastructure (TensorFlow, JAX/StableHLO, PyTorch Inductor, IREE), yet appears only in trace amounts in code-LM pretraining corpora. MLIR is also extensible by design: new d...
124. PEARL: Solver-in-the-Loop Interactive Optimization Modeling from Natural Language ​
Author: Hongliang Lu, Zhong Li, Yuxuan Chen, Yuan Lan, Fan Zhang, Zaiwen Wen
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.18256v1 Announce Type: cross Abstract: Optimization modeling is the process of translating real-world decision problems, often described in natural language, into formal mathematical formulations and executable solver code. While recent advances in large language models have shown promise...
125. MUX: Continuous Reasoning via Multiplexed Tokens ​
Author: Ayhan Suleymanzade, Halil Alperen Gozeten, Michael Bronstein, .Ismail .Ilkan Ceylan, Jinwoo Kim
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.18264v1 Announce Type: cross Abstract: Language models solve complex problems by articulating intermediate reasoning steps in natural language. While effective, this process is computationally bottlenecked: each reasoning step conveys only a single subword, and many are spent expressing a...
126. Wisdom of LLM Crowds: Aggregation and Contamination in Language Model Ensembles ​
Author: Igor Douven
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.18269v1 Announce Type: cross Abstract: The wisdom of crowds -- the finding that aggregating judgments across individuals often outperforms the best individual -- has been extensively studied with human forecasters. Whether the same phenomenon emerges when the ``crowd'' consists of large l...
127. Position: The Inevitable Transition to Machine Learning in Quantum Chemistry ​
Author: Karen Sargsyan, Chao-Ping Hsu
Published: 7/22/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, quant-ph
arXiv:2607.18281v2 Announce Type: cross Abstract: Finding exact solutions to the quantum many-body problem is computationally intractable (QMA-hard). Traditional approximations for electrons in an atom or molecule -- density functional theory and wavefunction methods -- have been indispensable, but ...
128. Disentangling Forced and Internal Climate Variability in Single Realizations using Dynamic Mode Decomposition with Control ​
Author: Nathan Mankovich, Andrei Gavrilov, Gustau Camps-Valls
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, physics.ao-ph
arXiv:2607.18298v1 Announce Type: cross Abstract: We show that a single climate realization can be decomposed into forced and internal components by treating external forcing as a dynamical driver within a linear stochastic system, an idea grounded in pullback attractor theory. In doing so, we addre...
129. On Incentivized Exploration beyond Bayesianism and Full-Information ​
Author: Dimitar Chakarov, Lee Cohen, Nathan Srebro
Published: 7/22/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2607.18300v1 Announce Type: cross Abstract: We extend Incentive Compatible Exploration beyond the Bayesian full-information setting of Kremer et al. [2014]. We consider agents that may possess external information unknown to the principal. We show such settings require new notions of incentivi...
130. Fretiq: Browser-Native Electric Guitar String Classification via Engineered Spectral Features and Held-Out Free-Play Evaluation ​
Author: Aadi Garg
Published: 7/22/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS
arXiv:2607.18303v1 Announce Type: cross Abstract: Identifying which string produces a given pitch in monophonic electric guitar audio is a fundamental classification challenge: a single pitch can often be produced on multiple strings at different fret positions, with timbral differences that prior l...
131. Approximating SPR Distance Between Phylogenetic Trees with Graph Neural Networks ​
Author: Renata Martins Castanheira, Miguel Bugalho, C'atia Vaz
Published: 7/22/2026, 4:00:00 AM
Categories: q-bio.PE, cs.AI, cs.LG
arXiv:2607.18311v1 Announce Type: cross Abstract: Comparing phylogenetic tree topologies is essential for understanding epidemic dynamics, yet biologically meaningful distances such as the Subtree Prune and Regraft (SPR) distance are NP-hard to compute and intractable on large datasets. We investiga...
132. EmoEUS: Uncertainty Supervision for Multimodal Emotion Recognition in Conversation ​
Author: Zilong Huang, Kong Aik Lee, Junjie Li, Zhe Li, Man-Wai Mak
Published: 7/22/2026, 4:00:00 AM
Categories: cs.MM, cs.CL, cs.LG
arXiv:2607.18336v1 Announce Type: cross Abstract: Multimodal emotion recognition in conversation (MERC) can leverage multimodal and contextual cues to boost recognition performance. However, existing fusion approaches in MERC often ignore modality-specific uncertainty across utterances caused by con...
133. PRISM: Sensitivity-Aware PolynoMial PRuning for EffIcient Neural Network Encryption ​
Author: Sahaj Majavdia, Mahdi Taheri
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.AR, cs.DC, cs.LG
arXiv:2607.18342v1 Announce Type: cross Abstract: Structured pruning is essential for making neural network inference feasible under homomorphic encryption (HE), yet its impact on model reliability has remained unexplored. This paper presents a systematic reliability characterization of pruned CKKS-...
134. Addressing Limited Data in Auditory Attention Decoding with Diffusion Generative Models ​
Author: David Rannaleet, Victor Gunnarsson, Bo Bernhardsson, Martin A. Skoglund, Emina Alickovic
Published: 7/22/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG
arXiv:2607.18345v1 Announce Type: cross Abstract: Limited training data constrains deep learning models for Auditory Attention Decoding (AAD) in hearing aids (HAs). AAD uses electroencephalogram (EEG) data to decode listener's attention, enabling real-time tracking of specific sound sources. However...
135. Decode-Time Grammars: Constrained LLM Generation over a Refinement Order of Grammar Fragments ​
Author: Shuoming Zhang, Ruiyuan Xu, Haofeng Li, Qiuchu Yu, Yangyu Zhang, Chunwei Xia, Xiaobing Feng, Chenxi Wang, Huimin Cui, Jiacheng Zhao
Published: 7/22/2026, 4:00:00 AM
Categories: cs.PL, cs.AI, cs.LG
arXiv:2607.18357v1 Announce Type: cross Abstract: Large language models now write a growing share of the world's code, increasingly inside agents and serving systems that compile, execute, or dispatch generated code without line-by-line review. This works well for mainstream languages but remains br...
136. A Classifier That Teaches Itself: Self-Improving, Frozen-gate Training (SIFT) for Dynamic Document Classification ​
Author: Bogdan Raduta, Horia Velicu, Alexandru Preda, Serban Chiricescu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.18358v1 Announce Type: cross Abstract: Document classification is a solved problem in the laboratory and an unsolved one in the enterprise. The blocker is rarely model architecture; it is the labeling project that must precede a model and the institutional fear of letting a model retrain ...
137. Decentralized Multi-agent Reinforcement Learning for Resilient Critical Infrastructures ​
Author: Minghui Ding, Evangelos Pournaras
Published: 7/22/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2607.18359v1 Announce Type: cross Abstract: Critical infrastructures are increasingly distributed, interdependent, and exposed to evolving disruptions, making resilience a central requirement for their operation and control. This paper argues that decentralized multi-agent reinforcement learni...
138. HALLMARK: Diagnosing Three Failure Modes in LLM Citation Verifiers ​
Author: Patrik Reizinger, Wieland Brendel
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.18360v1 Announce Type: cross Abstract: Large language models (LLMs) now routinely draft literature reviews and assist with academic writing, which means a higher risk of fabricated references: GPTZero found 53 papers with hallucinated citations among NeurIPS 2025's accepted set. Rule- and...
139. Adversarial Robustness of Phishing Email Detection: A Comparative Study of TF-IDF + Logistic Regression and Fine-Tuned DistilBERT ​
Author: Tanveer Ahmed, Seyedali Pourmoafil
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CY, cs.LG
arXiv:2607.18429v1 Announce Type: cross Abstract: Phishing emails remain one of the most persistent cybersecurity threats, and machine-learning classifiers are widely used to detect them. Most reported detection accuracies, however, are measured on clean, in-distribution test data rather than on ema...
140. Using binary silver labels in electronic health records-based computable phenotyping algorithms ​
Author: Shuhe Wang, Matthew T. Slaughter, Jennifer C. Nelson, Brian D. Williamson
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML
arXiv:2607.18431v1 Announce Type: cross Abstract: Gold-standard phenotype labels are often unavailable at scale in electronic health record (EHR) studies because they require manual chart review. Weakly supervised phenotyping methods instead use silver-standard labels, such as diagnosis-code counts,...
141. Relay-Bench: Evaluating LLMs on Multi-Domain Reasoning Chains ​
Author: Liam Swayne
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.18438v1 Announce Type: cross Abstract: Introducing Relay-Bench, an unsaturated, holistic, text-only benchmark that measures LLMs' ability to complete an assortment of tasks from distinct domains in a single prompt. The leading model, GPT-5.5 (xHigh), scores 43.3%. The test set entirely co...
142. Structured Output Collapses Answer Diversity Across 44 Language Models ​
Author: Tapan Parikh
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.18476v1 Announce Type: cross Abstract: When a language model must choose one answer from a large space of equally valid options, a format clause -- "Reply with JSON only" -- changes which answer it chooses. We re-run the One-Word Census (arXiv:2607.12796): 31 wide-answer-space category pr...
143. Recti-Q: Feature-Space Rectification for Out-of-Distribution-Robust Quantized Perception in Edge Robotics ​
Author: Hamidreza Yaghoubi Araghi, Parastoo Pilevar, Ming C. Lin
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO
arXiv:2607.18540v1 Announce Type: cross Abstract: Robotic perception pipelines increasingly rely on large vision backbones deployed on SWaP-constrained edge platforms, making post-training quantization (PTQ) attractive for real-time inference. However, while PTQ often preserves clean in-distribution...
144. Quantum Reservoir Computing: Recent Advances and Future Directions ​
Author: Shehbaz Tariq, Muhammad Talha, Arshid Ali, Muhammad Diyan, Symeon Chatzinotas
Published: 7/22/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2607.18552v1 Announce Type: cross Abstract: Quantum reservoir computing (QRC) uses the dynamics of a fixed or weakly tuned quantum system to transform temporal and sequential inputs into measured features, while training is typically confined to a classical readout. This separation reduces rel...
145. Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces ​
Author: Dongming Wang, Pengcheng Dai, Wenwu Yu, Wei Ren
Published: 7/22/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2607.18554v1 Announce Type: cross Abstract: We develop the Continuous Distributed Coupled Policy Gradient (CDCPG) algorithm for cooperative reinforcement learning in networked Markov decision processes with continuous state and action spaces. Each agent maintains a local actor over a bounded g...
146. Mixing-Free and Signal-Optimal Learning of Gaussian Graphical Models from Glauber Dynamics ​
Author: Vignesh Tirukkonda, Gautam Dasarathy
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2607.18559v1 Announce Type: cross Abstract: Gaussian graphical model selection is usually studied under independent sampling, but in many applications the data arise as a single trajectory of a dependent stochastic process. We study exact recovery of the graph from one trajectory of random-sca...
147. Attacking Graph Foundation Models Through Their Shared Representation ​
Author: Pankaj Kumar, Subhankar Mishra
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG
arXiv:2607.18567v1 Announce Type: cross Abstract: A graph foundation model generalizes across graph domains by mapping every input into one shared representation before any task reasoning. We call this map the alignment layer, the component that separates a graph foundation model from a graph neural...
148. GQD-AdsNet: Graph Neural Networks Unlock Rapid Exploration of Transition Metal Adsorption on Graphene Quantum Dots ​
Author: Lara Goncebat (Instituto de Qu'imica Aplicada del Litoral IQAL), Rodrigo Echeveste (Instituto de Investigaci'on en Se~nales, Sistemas e Inteligencia Computacional sinc), Mat'ias Gerard (Instituto de Investigaci'on en Se~nales, Sistemas e Inteligencia Computacional sinc), Frederik Tielens (General Chemistry), Gustavo Belletti (Instituto de Qu'imica Aplicada del Litoral IQAL), Paola Quaino (Instituto de Qu'imica Aplicada del Litoral IQAL)
Published: 7/22/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG
arXiv:2607.18591v1 Announce Type: cross Abstract: In recent years, interest in single-atom catalysts supported on carbon-based structures has grown considerably due to their high catalytic activity and efficient uses of metal atoms. However, the design and characterization of these materials through...
149. Intelligent Multi-UAV Navigation in ITNTNs: A Hierarchical LLM Approach ​
Author: Zijiang Yan, Hao Zhou, Wael Jaafar, Jianhua Pei, Ping Wang, Halim Yanikomeroglu, Hina Tabassum
Published: 7/22/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, cs.NI, cs.SY, eess.SY
arXiv:2607.18604v1 Announce Type: cross Abstract: The deployment of high-speed Uncrewed Aerial Vehicles (UAVs) in 3D aerial highways necessitates robust coordination of physical flight kinematics and multi-tier network handovers. While Deep Reinforcement Learning (DRL) offers rapid tactical control,...
150. Stochastic Meta-Unlearning: Bridging Language Backbone and Multimodal Unlearning ​
Author: Zijie Liu, Jinhao Duan, Gaowen Liu, Sijia Liu, Tianlong Chen
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.18615v1 Announce Type: cross Abstract: Machine unlearning for vision-language models (VLMs) remains underexplored. Unlike language models, VLMs combine a language backbone with visual components, which makes unlearning more complex. There is a surprising phenomenon when moving from single...
151. LatentMT: Machine Translation with Latent Reasoning ​
Author: Wei-Rui Chen, Samar M. Magdy, Chiyu Zhang, Wenhui Zhu, Zhipeng Wang, Muhammad Abdul-Mageed
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.18618v1 Announce Type: cross Abstract: Latent-reasoning looped language models (LoopLMs) offer a different scaling path for machine translation (MT): instead of increasing parameter count or emitting explicit chain-of-thought tokens, they spend additional recurrent computation inside hidd...
152. End-to-end Conditional Diffusion for Realistic and Controllable Visual Traffic Scenario Generation ​
Author: Jingzheng Li, Yufei Ge, Zhijun Chen, Qianren Mao, Zizhe Wang, Binhang Qi, Bing Li, Keyu Chen, Baochang Zhang, Xianglong Liu, Philip S Yu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.18637v1 Announce Type: cross Abstract: Generating closed-loop traffic scenarios that are both realistic and controllable is crucial for evaluating autonomous driving systems, especially under rare safety-critical interactions. Existing learning-based methods often struggle to balance cont...
153. The Price of Hidden Curvature: An $\widetilde{\Omega} (d^{5/4} \sqrt{T})$ Lower Bound for Bandit Convex Optimization ​
Author: Nived Rajaraman
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT
arXiv:2607.18652v1 Announce Type: cross Abstract: We establish a $\widetilde\Omega(d^{5/4}\sqrt T)$ lower bound on the minimax expected regret of stochastic bandit convex optimization of $1$-Lipschitz functions on the Euclidean ball. This presents the first nontrivial regret lower bound that grows f...
154. Cross-Dataset Generalization in Breast MRI Tumor Classification via Class-Wise Dataset Mixing ​
Author: Mohammad Ali Dadrast, Hamid Usefi
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.18678v1 Announce Type: cross Abstract: Breast MRI is highly sensitive for detecting breast tumors, but exams contain many slices and require substantial reading time. Deep learning models often perform well on internal splits but can fail across institutions because of domain shift and da...
155. Attributes Should Come from Images, Not Class Names: Distribution-Conditioned Attribute Selection for Vision-Language Models ​
Author: Gautam Rajendrakumar Gare, Jia Shi, Zhiqiu Lin, Deepak Pathak, John Galeotti, Deva Ramanan
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV
arXiv:2607.18695v1 Announce Type: cross Abstract: A popular route to interpretable zero-shot classification asks a large language model (LLM) to describe each class name and prompts CLIP with the resulting descriptors. We show that these descriptors carry little visual evidence of their own: removin...
156. Decoupled Pipeline with Proposal Reranking and Score Fusion for Positive-Unlabeled Marine Species Detection ​
Author: Robert James Brock, Sebastian Maximilian Krupa, Jason Kahei Tam
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.18700v1 Announce Type: cross Abstract: The FathomNetCLEF 2026 competition combines underwater object detection and fine-grained marine species classification under a positive-unlabeled evaluation setting. The provided training labels are sparse, while the hidden test set is out-of-distrib...
157. Algebraic Signatures for Structural Learning in Probability Tensors ​
Author: Akihiro Maeda, Shohei Hidaka, Satoshi Aoki
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.18817v1 Announce Type: cross Abstract: Algebraic statistics characterizes statistical models through polynomial constraints, but it has mainly been used for analytically specified model classes. This paper studies the inverse problem: identifying probabilistic structure from vanishing bin...
158. NSMA: Neuro-Symbolic Manifold Alignment for Generalizable Adaptive Bitrate Streaming under Texture Shift ​
Author: Zhiqiang He, Zhi Liu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.NI, cs.LG
arXiv:2607.18845v1 Announce Type: cross Abstract: For decades, ABR has kept two kinds of intelligence apart. Neural policies learn rich behaviors yet forget them the moment the environment changes; rules never learn, and never forget. Every prior attempt to combine them has kept this separation, let...
159. Enhanced Neural Quantum State via Annealed Gradient Descent ​
Author: Shiwei Zhou, Yiming Huang, Xiao Yuan, Xiaoxia Cai
Published: 7/22/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, physics.chem-ph, physics.comp-ph
arXiv:2607.18865v2 Announce Type: cross Abstract: Neural quantum states offer expressive representations of quantum many-body wave functions, yet their practical accuracy can be limited by stochastic optimization rather than representational capacity. Here we identify a finite-sample instability, te...
160. Optimizing Regret ​
Author: Irene Aldridge
Published: 7/22/2026, 4:00:00 AM
Categories: econ.EM, cs.LG, stat.ML
arXiv:2607.18866v1 Announce Type: cross Abstract: Building on the identity that expected regret equals the covariance between costs and decisions, this paper develops the complete derivative theory of the covariance regret functional. We derive the G^ateaux derivative, showing that the universal st...
161. Local Label-Informed Feature Transfer for Generating Ground-Truth Medical Images: A Comparison of GAN- and Diffusion-Based Approaches ​
Author: Rick Wilming, Irem Ozseker, Luca Matteo Cornils, Ahc`ene Boubekki, Benedict Clark, Danny Panknin, Stefan Haufe
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.18882v1 Announce Type: cross Abstract: Validating Explainable Artificial Intelligence (XAI) methods in medical imaging requires ground-truth data with known locations of informative features. However, current approaches rely on expert annotations, which are prone to labeling errors, or on...
162. Measuring Reward-Seeking via Contrastive Belief Updates ​
Author: Axel H{\o}jmark, J'er'emy Scheurer, Evgenia Nitishinskaya, Felix Hofst"atter, Jason Wolfe, Theodore Ehrenborg, Bronson Schoen, Alexander Meinke
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.18966v1 Announce Type: cross Abstract: Language models trained with reinforcement learning may learn to optimize the grader's judgment rather than the intended objective. This "reward-seeking" is difficult to measure because a model that pursues the grader's judgment and one that pursues ...
163. Verifiable Self-Evolution for Open-Ended Dialogue Skills via Future-Feedback Prediction ​
Author: ChaoJin Zhao, Xuan Jiang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.18973v1 Announce Type: cross Abstract: Textual skills provide a lightweight way to improve frozen language-model agents, but their self-evolution normally requires a stable validation signal. Such signals are natural in mathematics or code, where an answer can be checked after it changes,...
164. Benchmarking Deep Learning Approaches for AEC Engineering Drawing Layout Detection and Information Extraction ​
Author: Tianyang Huang, Alessio Lombardi, Ahmed Elnagar, Ahmed Zalouk, George Paul, Sepehr Najjarpour, Arvid Sigurdsson, Khalid Ismail, Mohamed Ragab, Edlira Vakaj
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.CE, cs.LG
arXiv:2607.18997v1 Announce Type: cross Abstract: Information Extraction (IE) from Architecture, Engineering, and Construction (AEC) drawings remains hindered by manual inefficiency, while Layout Detection, a vital 'middleware' organizing graphical and textual hierarchies, is underexplored. General ...
165. The Tractability Landscape of Sampling with Inexact Scores ​
Author: Anming Gu, Kevin Tian, Hubert Yang, Yusong Zhu
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2607.19004v1 Announce Type: cross Abstract: We provide a simple and tight characterization of the types of inexact score oracle access that permit sampling with vanishing total variation bias, for a standard, well-behaved target family. Our main result shows that any weaker error than the sub-...
166. Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing ​
Author: Xinjie Zhang, Peng Zhang, Shicheng Zheng, Jinghao Guo, Zhaoyang Jia, Yifei Shen, Xun Guo, Yuxuan Luo, Jiahao Li, Wenxuan Xie, Fanyi Pu, Xiaoyi Zhang, Kaichen Zhang, Zongyu Guo, Tianci Bi, Dongnan Gui, Zhening Liu, Zimo Wen, Zihan Zheng, Senqiao Yang, Xiao Li, Jinglu Wang, Bin Li, Yan Lu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MM, eess.IV
arXiv:2607.19064v2 Announce Type: cross Abstract: Large-scale visual generators are increasingly capable but costly to train, fine-tune, and deploy. We introduce Mage-Flow, a compact 4B-scale generative stack for efficient text-to-image generation and instruction-based image editing. The stack is bu...
167. Translation as Augmentation: Effect of Translated Data on Assessment of Difficulty ​
Author: Yiheng Wu, Jue Hou, Roman Yangarber
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.19101v1 Announce Type: cross Abstract: Reliable Text Difficulty Assessment is a prerequisite for valid text simplification workflows and personalized learning applications. However, the development of robust assessment models is severely hindered by a critical bottleneck: the scarcity of ...
168. MIRAGE: Multi-scale Lesion-Informed Representation with Auxiliary Guidance for MRI Contrast Enhancement ​
Author: Andrea Borghesi, Xin Wang, Jonas Teuwen, George Yiasemis
Published: 7/22/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.LG
arXiv:2607.19137v1 Announce Type: cross Abstract: Inferring contrast enhancement from one pre-contrast breast MRI slice is underdetermined: post-contrast appearance contains physiological information that is not uniquely encoded in baseline anatomy. Optimizing only paired pixel fidelity can suppress...
169. Boundary-Adapted PINNs for Elliptic Dirichlet Problems: $H^2(\Omega)$ A Priori Error Bounds with Application to Mean Escape Time Computation ​
Author: Nathanael Tepakbong, Jun Fan, Xiang Zhou, Ding-Xuan Zhou
Published: 7/22/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.ST, stat.ML, stat.TH
arXiv:2607.19167v1 Announce Type: cross Abstract: Motivated by the numerical computation of the Mean Escape Time (MET) $\tau:\Omega\to\mathbb{R}$ of a stochastic process from a bounded domain $\Omega\subseteq\mathbb{R}^d$, we study elliptic Dirichlet boundary value problems (BVPs) using boundary-enf...
170. Reasoning Before Translation: Enhancing Legal Machine Translation with Structured Reasoning ​
Author: Aixiu An, Michael Jungo, Eloi Eynard, Mark Drenhaus, Andreas Fischer, Jean Hennebert, S'ebastien Rumley
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.19181v1 Announce Type: cross Abstract: Neural machine translation (NMT) in the legal domain is a linguistically and conceptually demanding task, primarily due to the complexity of legal language and the high level of precision it requires. The recent emergence of reasoning-capable languag...
171. ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU ​
Author: Fan Jiang, Zhaoxu Sun, Mengchao Wang, Ziyu Zhu, Chiyu Wang, Yunpeng Zhang, Wenlin Liu, Yun Wang, Xue Zheng, Rui Sun, Junfeng Ni, Hongyu Pan, Zhongxu Sun, Fei Yu, Zengye Ge, Mengmeng Du, Nianfei Fan, Mingchao Sun, Yu Liu, Yongchang, Yanqing Zhu, Jiahang Wang, Ning Ying, Yuze Xuan, Di Yang, Zhicheng Liu, Zhe Gao, Tingbing Xu, Jiacheng Sui, Wenjin Yang, Junnan Lai, Shufeng Liu, Yuan Liu, Zheng Zhou, Yingliang Peng, Dawei Cao, Kaifeng Sheng, Yuxiang Cai, Fei Lu, Mu Xu, Ning Guo
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.19191v1 Announce Type: cross Abstract: We present ABot-World-0, an action-conditioned video world model for real-time, long-horizon closed-loop interaction, supported by a multi-source data infrastructure spanning AAA games, simulation engines, and internet videos to learn controllable wo...
172. ATLAS: A Foundation Neural Sampler for Amorphous Materials ​
Author: Mouyang Cheng, Denis Blessing, Botao Yu, Gerhard Neumann, Mingda Li, Carles Domingo-Enrich, Yuanqi Du
Published: 7/22/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG, physics.comp-ph
arXiv:2607.19198v1 Announce Type: cross Abstract: Amorphous materials exhibit exceptional mechanical and functional properties, yet their rugged energy landscapes are notoriously difficult to sample. Below the glass-transition temperature, conventional molecular dynamics and Monte Carlo become ineff...
173. Assessment in Team Problem-Solving Exercises in Computing Education ​
Author: Valdemar \v{S}v'abensk'y, Jan Vykopal, Sukrit Leelaluk, Pavel \v{C}eleda, Fumiya Okubo, Atsushi Shimada
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG
arXiv:2607.19209v1 Announce Type: cross Abstract: This full paper in the research-to-practice track presents methods for assessing student teams in tabletop exercises (TTXs). TTXs enable learner teams to prepare for workplace tasks and practice crisis responses, such as resolving cybersecurity incid...
174. The Price of Reasoning: Cost-Quality Tradeoffs in Reinforcement Learning for Neural Machine Translation ​
Author: Michael Jungo, Aixiu An
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.19226v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has been established as a viable paradigm for the post-training of Large Language Models (LLMs), including downstream tasks, such as Neural Machine Translation (NMT). With the latest research indi...
175. ResearchArena: Evaluating Sabotage and Monitoring in Automated AI R&D ​
Author: Lena Libon, Ben Rank, Jehyeok Yeon, David Schmotz, Jeremy Qin, Daniel Donnelly, Derck Prinzhorn, Maksym Andriushchenko
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG
arXiv:2607.19321v1 Announce Type: cross Abstract: As AI agents begin to automate AI R&D, we need ways to assess whether their outputs are safe to deploy, even when the agents themselves may be untrusted. AI control offers one such approach: rather than trusting the agent, it treats it as a potential...
176. Fundamental limits of distributed multiclass classification from simple binary decisions ​
Author: Ioannis Papageorgiou, Srinivas Nomula, Ayalvadi Ganesh, Sidharth Jaggi, Parimal Parag
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT, math.ST, stat.TH
arXiv:2607.19334v1 Announce Type: cross Abstract: We consider the problem of constructing a $K$-class classifier from the combination of $O(\log K)$ simple binary classifiers -- this is a natural paradigm to construct a sophisticated classifier in a distributed manner with each agent performing a re...
177. 1-Lipschitz Neural Networks on Hadamard Manifolds ​
Author: Davide Murari, Marta Ghirardelli, Ben Adcock, Elena Celledoni, Brynjulf Owren, Carola-Bibiane Sch"onlieb
Published: 7/22/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA
arXiv:2607.19335v1 Announce Type: cross Abstract: Controlling the Lipschitz constant of a neural network is a standard way to promote robustness and stability. Most existing constraining strategies are designed for Euclidean spaces. In this work, we construct and analyze a class of 1-Lipschitz neura...
178. Soft-TransFormers for Continual Learning ​
Author: Haeyong Kang, Chang D. Yoo
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2411.16073v4 Announce Type: replace Abstract: Inspired by the Well-initialized Lottery Ticket Hypothesis (WLTH), we introduce Soft-TransFormers (Soft-TF), a continual learning framework that adapts a frozen pre-trained Transformer through task-specific soft subnetworks: real-valued multiplicat...
179. A Self-Supervised Framework for Space Object Behaviour Characterisation ​
Author: Ian Groves, Andrew Campbell, James Fernandes, Diego Ram'irez Rodr'iguez, Paul Murray, Massimiliano Vasile, Victoria Nockles
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, physics.space-ph
arXiv:2504.06176v4 Announce Type: replace Abstract: Foundation Models, which leverage large neural networks pre-trained on unlabelled data before fine-tuning for specific tasks, are increasingly being applied to specialised domains. Recent examples include ClimaX for climate and Clay for satellite E...
180. Parameter-Efficient Continual Fine-Tuning: A Survey ​
Author: Eric Nuertey Coleman, Luigi Quarantiello, Ziyue Liu, Qinwen Yang, Samrat Mukherjee, Julio Hurtado, Vincenzo Lomonaco
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2504.13822v3 Announce Type: replace Abstract: The emergence of large pre-trained networks has revolutionized the AI field, unlocking new possibilities and achieving unprecedented performance. However, these models inherit a fundamental limitation from traditional Machine Learning approaches: t...
181. Topology-Driven Clustering: Enhancing Performance with Betti Number Filtration ​
Author: Arghya Pratihar, Kushal Bose, Swagatam Das
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2505.04346v2 Announce Type: replace Abstract: Clustering aims at partitioning data points into groups of similar objects without knowing about the class labels. However, clustering datasets with complex geometric structures, such as nonconvex shapes, multiple scales, or intertwined manifolds, ...
182. Potential failures of physics-informed machine learning in traffic flow modeling: theoretical and experimental analysis ​
Author: Yuan-Zheng Lei, Yaobang Gong, Dianwei Chen, Yao Cheng, Xianfeng Terry Yang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2505.11491v4 Announce Type: replace Abstract: This study investigates why physics-informed machine learning (PIML) can fail in macroscopic traffic flow modeling. We define failure as cases where a PIML model underperforms both purely data-driven and purely physics-based baselines by a given th...
183. Chi-Square Wavelet Graph Neural Networks for Heterogeneous Graph Anomaly Detection ​
Author: Xiping Li, Xiangyu Dong, Xingyi Zhang, Kun Xie, Yuanhao Feng, Bo Wang, Guilin Li, Wuxiong Zeng, Xiujun Shu, Sibo Wang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR, cs.SI
arXiv:2505.18934v2 Announce Type: replace Abstract: Graph Anomaly Detection (GAD) in heterogeneous networks presents unique challenges due to node and edge heterogeneity. Existing Graph Neural Network (GNN) methods primarily focus on homogeneous GAD and thus fail to address three key issues: (C1) Ca...
184. How Benchmark Prediction from Fewer Data Misses the Mark ​
Author: Guanhua Zhang, Florian E. Dorner, Moritz Hardt
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2506.07673v2 Announce Type: replace Abstract: Large language model (LLM) evaluation is increasingly costly, prompting interest in methods that speed up evaluation by shrinking benchmark datasets. Benchmark prediction (also called efficient LLM evaluation) aims to select a small subset of evalu...
185. Can Interpretation Predict Behavior on Unseen Data? ​
Author: Victoria R. Li, Jenny Kaufmann, Tian Qin, Martin Wattenberg, David Alvarez-Melis, Naomi Saphra
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2507.06445v3 Announce Type: replace Abstract: Interpretability research often predicts model responses to targeted mechanistic interventions. But can we predict responses to unseen input data? We propose and demonstrate this alternate objective by using model internals to predict their out-of-...
186. Onboarding Without Forgetting: Hypernetwork Personalization with Data-Free Replay for Personalized Federated Learning ​
Author: Thinh Nguyen, Le Huy Khiem, Van-Tuan Tran, Khoa D Doan, Nitesh V Chawla, Kok-Seng Wong
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2508.05157v2 Announce Type: replace Abstract: Federated Learning (FL) enables collaborative training across distributed clients without sharing raw data, offering strong privacy benefits. However, most methods assume all clients remain available throughout training, which is unrealistic as new...
187. HERAKLES: Hierarchical Skill Compilation for Open-ended LLM Agents ​
Author: Thomas Carta, Cl'ement Romac, Loris Gaven, Pierre-Yves Oudeyer, Olivier Sigaud, Sylvain Lamprier
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2508.14751v2 Announce Type: replace Abstract: We study goal-conditioned reinforcement learning in partially observable environments with sparse rewards and large, structured goal spaces. In such settings, complex goals often require composing simpler skills, but learning these compositions eff...
188. Finite-Agent Stochastic Differential Games on Large Graphs: II. Graph-Based Architectures ​
Author: Ruimeng Hu, Jihao Long, Haosheng Zhou
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.GT, math.OC
arXiv:2509.12484v2 Announce Type: replace Abstract: We propose a novel neural network architecture, called Non-Trainable Modification (NTM), for computing Nash equilibria in stochastic differential games (SDGs) on graphs. These games model a broad class of graph-structured multi-agent systems arisin...
189. Budgeted Indirect Adversarial Attack on Graph-Based Anomaly Detection in Sensor Networks ​
Author: Sanju Xaviar, Omid Ardakanian
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.17987v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have emerged as powerful models for anomaly detection in sensor networks, particularly when analyzing multivariate time series. In this work, we introduce BETA, a novel indirect evasion attack targeting such GNN-based d...
190. Multi-Agent Inverted Transformer for Flight Trajectory Prediction ​
Author: Seokbin Yoon, Keumjin Lee
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.21004v3 Announce Type: replace Abstract: Flight trajectory prediction for multiple aircraft is essential and provides critical insights into how aircraft navigate within current air traffic flows. However, predicting multi-agent flight trajectories is inherently challenging. One of the ma...
191. SHUFFLESPARSE: Learned Shuffles for Structured Sparse Networks ​
Author: Abhishek Tyagi, Arjun Iyer, Liam Young, William H Renninger, Christopher Kanan, Yuhao Zhu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.14812v2 Announce Type: replace Abstract: Structured weight sparsity accelerates training and inference on modern GPUs, but it trails unstructured dynamic sparse training (DST) in accuracy especially at extreme sparsity. We pinpoint the reason for this difference in performance to a lack o...
192. Data Reliability Scoring ​
Author: Yiling Chen, Shi Feng, Paul Kattuman, Fang-Yi Yu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.GT, stat.ML
arXiv:2510.17085v2 Announce Type: replace Abstract: How can we assess the reliability of a dataset without access to ground truth? We introduce the problem of reliability scoring for datasets collected from potentially strategic sources. The true data are unobserved, but we see outcomes of an unknow...
193. Hierarchical Physics-Embedded Learning for Partially Known Spatiotemporal Dynamics ​
Author: Xizhe Wang, Xiaobin Song, Hongbo Zhao, Qingshan Jia, Qianchuan Zhao, Hao Sun, Benben Jiang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.25306v3 Announce Type: replace Abstract: Partial physical knowledge--governing structures known, constitutive relations or their combinations not--pervades spatiotemporal systems. Existing scientific machine learning paradigms learn evolution largely from data, impose equations as soft co...
194. Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning ​
Author: Minh Vu, Konstantinos Slavakis
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.18763v2 Announce Type: replace Abstract: Unlike their conventional use as estimators of probability density functions in reinforcement learning (RL), this paper introduces a novel function-approximation role for Gaussian mixture models (GMMs) as direct surrogates for Q-function losses. Th...
195. Toward Learning POMDPs Beyond Full-Rank Actions and State Observability ​
Author: Seiji Shaw, Travis Manderson, Chad Kessens, Nicholas Roy
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO
arXiv:2601.18930v4 Announce Type: replace Abstract: We are interested in enabling autonomous agents to learn and reason about systems with hidden states, such as locking mechanisms. We cast this problem as learning the parameters of a discrete Partially Observable Markov Decision Process (POMDP). Th...
196. Automatic Construction of Clinical Scoring Systems with LLM Agents ​
Author: Silas Ruhrberg Est'evez, Christopher Chiu, Mihaela van der Schaar
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.MA
arXiv:2601.22324v3 Announce Type: replace Abstract: Modern clinical practice relies on evidence-based guidelines implemented as compact scoring systems composed of a small number of interpretable decision rules. While machine-learning models achieve strong performance, many fail to translate into ro...
197. Knowledge-Informed Kernel State Reconstruction from Heterogeneous Partial Observations ​
Author: Luca Muscarnera, Silas Ruhrberg Est'evez, Samuel Holt, Evgeny Saveliev, Mihaela van der Schaar
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.22328v3 Announce Type: replace Abstract: Real-world scientific systems are rarely observed through complete, regularly sampled state trajectories. Instead, measurements are often partial, noisy, and heterogeneous, providing fragmented views of latent dynamical states. We introduce MAAT (M...
198. CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation ​
Author: Ning Yang, Chengzhi Wang, Yibo Liu, Baoliang Tian, Haijun Zhang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.08686v3 Announce Type: replace Abstract: Prefill-only KV compression freezes a token subset at the end of prefill and decodes from it without further eviction. The retention decision is therefore irreversible, yet existing methods estimate the corrective signals it relies on, per-head rel...
199. Toward Manifest Relationality in Transformers via Symmetry Reduction ​
Author: J. Fran\c{c}ois, L. Ravera
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, hep-th, stat.ML
arXiv:2602.18948v2 Announce Type: replace Abstract: Transformer models contain substantial internal redundancy arising from coordinate-dependent representations and continuous symmetries, in model space and in head space, respectively. While recent approaches address this by explicitly breaking symm...
200. Discrete Diffusion with Sample-Efficient Estimators for Conditionals ​
Author: Karthik Elamvazhuthi, Abhijith Jayakumar, Andrey Y. Lokhov
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2602.20293v3 Announce Type: replace Abstract: We study a discrete denoising diffusion framework that integrates a sample-efficient estimator of single-site conditionals with round-robin noising and denoising dynamics for generative modeling over discrete state spaces. Rather than approximating...
201. Information Theoretic Bayesian Optimization over the Probability Simplex ​
Author: Federico Pavesi, Antonio Candelieri, No'emie Jaquier
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.09793v2 Announce Type: replace Abstract: Bayesian optimization is a data-efficient technique that has been shown to be extremely powerful to optimize expensive, black-box, and possibly noisy objective functions. Many applications involve optimizing probabilities and mixtures which natural...
202. CLT-Forge: A Scalable Library for Cross-Layer Transcoders and Attribution Graphs ​
Author: Florent Draye, Vedant Palit, Abir Harrasse, Tung-Yu Wu, Jiarui Liu, Punya Syon Pandey, Roderick Wu, Chih-Hao Hsu, Terry Jingchen Zhang, Zhijing Jin, Bernhard Sch"olkopf
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2603.21014v2 Announce Type: replace Abstract: Mechanistic interpretability seeks to understand how Large Language Models (LLMs) represent and process information. Recent approaches based on dictionary learning and transcoders enable representing model computation in terms of sparse, interpreta...
203. Robust Reasoning Benchmark ​
Author: Pavel Golikov, Evgenii Opryshko, Gennady Pekhimenko, Mark C. Jeffrey
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2604.08571v3 Announce Type: replace Abstract: While Large Language Models (LLMs) achieve high performance on standard mathematical benchmarks, their problem-solving abilities depend on the context and textual formatting. We introduce the Robust Reasoning Benchmark (RRB), a pipeline of 13 deter...
204. Geometric Capacity of Transformers: A Tropical Geometry Perspective ​
Author: Ye Su, Yong Liu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.14727v2 Announce Type: replace Abstract: To quantify the geometric capacity of transformers, we develop a tropical-geometric framework for analyzing the spatial partitions induced by conditioned self-attention. In the zero-temperature limit, we show that fixed-key top-$1$ routing is exact...
205. Understanding Self-Supervised Learning via Latent Distribution Matching ​
Author: Fabian A Mikulasch, Friedemann Zenke
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2605.03517v4 Announce Type: replace Abstract: Self-supervised learning (SSL) excels at finding general-purpose latent representations from complex data, yet lacks a unifying theoretical framework that explains the diverse existing methods and guides the design of new ones. We cast SSL as laten...
206. How Many Iterations to Jailbreak? Dynamic Budget Allocation for Multi-Turn LLM Evaluation ​
Author: Shai Feldman, Yaniv Romano
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.06605v5 Announce Type: replace Abstract: Evaluating and predicting the performance of large language models (LLMs) in multi-turn conversational settings is critical yet computationally expensive; key events -- e.g., jailbreaks or successful task completion by an agent -- often emerge only...
207. More Edits, More Stable: Understanding the Lifelong Normalization in Sequential Model Editing ​
Author: Xin Ma, Wei Chen, Qi Liu, Derong Xu, Zhi Zheng, Tong Xu, Enhong Chen
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2605.11836v2 Announce Type: replace Abstract: Lifelong Model Editing aims to continuously update evolving facts in Large Language Models while preserving unrelated knowledge and general capabilities, yet it remains plagued by catastrophic forgetting and model collapse. Empirically, we find tha...
208. GQLA: Group-Query Latent Attention for Hardware-Adaptive Large Language Model Decoding ​
Author: Fanxu Meng
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.15250v3 Announce Type: replace Abstract: Multi-head Latent Attention (MLA), the attention used in DeepSeek-V2/V3, jointly compresses keys and values into a low-rank latent and matches the H100 roofline almost perfectly. Its trained weights, however, expose only one decoding path - an abso...
209. Tunable MAGMAX: Preference-Aware Model Merging for Continual Learning ​
Author: Kei Hiroshima, Kento Uchida, Shinichi Shirakawa
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.20803v3 Announce Type: replace Abstract: Continual learning (CL) aims to train models sequentially on multiple tasks while mitigating catastrophic forgetting of previously learned knowledge. Recent advances in large pre-trained models (LPMs) and model merging techniques, such as MAGMAX, h...
210. Partial Fusion of Neural Networks: Efficient Tradeoffs Between Ensembles and Weight Aggregation ​
Author: Fabian Morelli, Stephan Eckstein
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2605.22350v2 Announce Type: replace Abstract: Ensembles of neural networks typically outperform individual networks but incur large computational costs, whereas weight aggregation produces less costly, yet also less accurate, aggregate models. We introduce partial fusion of networks, which int...
211. Multi$^2$: Hierarchical Multi-Agent Decision-Making with LLM-Based Agents in Interactive Environments ​
Author: Sangeun Park, Minhae Kwon
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.03698v2 Announce Type: replace Abstract: A central goal of large language model (LLM) research is to build agentic systems that can plan, act, and adapt through sustained interaction with dynamic environments. While recent LLM-based agents exhibit impressive contextual reasoning, their lo...
212. Phantoms and Disclosures: A Statistical Framework for Auditing Privacy in Synthetic Data ​
Author: Kareem Amin, Rudrajit Das, Alessandro Epasto, Adel Javanmard, Dennis Kraft, M'onica Ribero, Sergei Vassilvitskii
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.AP, stat.ME, stat.ML
arXiv:2606.16952v2 Announce Type: replace Abstract: The rapid adoption of generative AI and Large Language Models (LLMs) has spurred interest in synthetic data as a privacy-preserving alternative to sensitive real-world datasets. However, generating high-utility synthetic data often carries the risk...
213. ReRAM-aware Model Finetuning addressing I-V Non-linearity and Retention Errors ​
Author: Ching-Yi Lin, Shamik Kundu, Arnab Raha, Sahil Shah
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2606.17471v3 Announce Type: replace Abstract: Traditional CPU, GPU, and NPU architectures are increasingly limited by the von Neumann bottleneck. While In-Memory Computing (IMC) using ReRAM crossbar arrays offers a high-density, energy-efficient alternative, its practical deployment is constra...
214. Sign-Rank, Index, and List Replicability: Connections and Separations ​
Author: Ari Blondal, Hamed Hatami, Pooya Hatami, Chavdar Lalov, Sivan Tretiak
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2606.18236v2 Announce Type: replace Abstract: In learning theory, the sign-rank of a binary concept class captures the smallest dimension in which it can be represented by points and halfspaces. Despite tremendous interest, lower bounds on sign-rank are notoriously difficult to come by. Two re...
215. Is Variational Monte Carlo Robust? Sharp Moment Thresholds and Heavy-tailed Stochastic Optimization ​
Author: Philipp Grohs, Davide Nobile
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.26009v2 Announce Type: replace Abstract: Variational Monte Carlo (VMC) is a central algorithm in electronic structure theory and has gained renewed importance through modern neural-network ans"atze such as FermiNet. At its core, VMC seeks ground states by minimizing the Rayleigh quotient...
216. A Transdiagnostic Space of Disorder Like Phenotypes in Reinforcement Learning Agents ​
Author: Hari Prasad
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.07753v2 Announce Type: replace Abstract: Modelling psychological disorders in artificial agents offers a testbed for computational psychiatry and a lens on affective-control failure modes. Prior work induces one or two disorders by hand-tuned reward shaping, labels the behaviour post hoc,...
217. LieBN: Batch Normalization over Lie Groups ​
Author: Ziheng Chen, Yue Song, Rui Wang, Xiao-Jun Wu, Nicu Sebe
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.08783v3 Announce Type: replace Abstract: Manifold-valued measurements are prevalent in various machine learning tasks. Recent advances have extended Deep Neural Networks (DNNs) to operate on manifolds. These extensions have been accompanied by normalization techniques tailored to differen...
218. RUBRIC: Realism--Utility Balanced Ranking for Imbalanced Classification ​
Author: Yanxuan Yu, Dong Liu, Eric Jiang, Shu Wang, Wenxiao Zhao, Jinxi Yu, Shaoyi Lu, Hui Pan, Renata Borovica-Gajic, Ying Nian Wu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.09816v2 Announce Type: replace Abstract: Class imbalance poses a fundamental challenge in risk-sensitive applications such as fraud detection and medical diagnosis, where minority-class samples are scarce yet critical for accurate classification. Existing oversampling methods generate syn...
219. ARMOR: Stabilizing On-Policy LLM RL with Off-Policy Anchor Samples ​
Author: Kexin Huang, Junkang Wu, Jinda Lu, Shuo Yang, Chiyu Ma, Jiancan Wu, Xiang Wang, Xiangnan He, Guoyin Wang, Jingren Zhou
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.10481v2 Announce Type: replace Abstract: Reinforcement learning (RL) has significantly enhanced the reasoning capabilities of large language models (LLMs), yet the training process remains notoriously fragile. In this work, we investigate a critical source of this instability: over-optimi...
220. Quantum Port-Hamiltonian Neural Networks: Learning Conservative and Dissipative Dynamics via Measurement-Induced Nonlinearity ​
Author: Dibakar Sigdel
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.12269v2 Announce Type: replace Abstract: We introduce Quantum Port-Hamiltonian Neural Networks (Q-pHNNs), a family of parameterised quantum circuits that learn classical dynamics in a structure-preserving manner. The framework relies on the Isomorphic Hamiltonian Mapping (IHM): the skew-s...
221. A VAE-Driven Multi-Task Satellite-Aided Semantic Communication Framework for 6G-Enabled Connected Autonomous Vehicles ​
Author: S. M. Abtahiul Alam, Niloy Das, Apurba Adhikary, Yu Qiao, Zhu Han, Choong Seon Hong
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2607.13494v2 Announce Type: replace Abstract: The development of smart transportation systems and the introduction of 6G wireless communication technologies have significantly changed vehicle network topologies. Future connected autonomous vehicle (CAV) networks require bandwidth-efficient, re...
222. Long-term User Engagement Optimization through Model-agnostic Downstream Rewards Learning ​
Author: Dingsu Wang, Filip Ryzner, Kelly He, Armando Ordorica, David Woo, Aditya Mantha, Liyao Lu, Usha Amrutha Nookala, Haoran Guo, Jiacong He, Olafur Gudmundsson, Matt Chun, Krystal Benitez, Dhruvil Deven Badani, Yijie Dylan Wang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2607.14192v2 Announce Type: replace Abstract: As recommender systems mature in the past few years, their optimization objectives have evolved from a primary focusing on short-term behavioral signals to a broader emphasis on long-term user engagement and retention. However, directly optimizing ...
223. Operator-Aware Mixed-Precision Tolerance Calibration for Tensor Kernels ​
Author: Dipankar Sarkar
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.MS
arXiv:2607.16228v2 Announce Type: replace Abstract: Most tensor-kernel correctness tests go through a fixed-shape all close-style check with hand-picked absolute and relative tolerances. The thresholds are copied across the corpus and rarely revisited. We mine the element-wise error distribution of ...
224. OpenMHC: Accelerating the Science of Wearable Foundation Models ​
Author: Narayan Schuetz, Yuze Bai, Lianggang Pan, Edgar Eggert, Favour Nerrise, Juan Delgado-SanMartin, Max Rosenblattl, Milana Gurbanova, Mohammad Asadi, Anders Johnson, Paul Schmiedmayer, Dennis Wang, Allan Lawrie, Daniel Seung Kim, Xin Liu, Akshay Paruchuri, Ehsan Adeli, Euan Ashley, Kelly W. Zhang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16235v2 Announce Type: replace Abstract: Mobile and wearable devices offer an unprecedented opportunity for continuous, passive health monitoring and active health coaching. However, the largest wearable datasets are not publicly available for research, and leading wearable foundation mod...
225. Discovery by Dreaming: Cross-Domain Recombination in Artificial Memory ​
Author: Oliver Zahn, James Evans, David Eagleman
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR, cs.NE
arXiv:2607.16256v2 Announce Type: replace Abstract: Dreams splice together people, places, and times that never met. Neuroscience suggests this recombination is not noise, but a function driving insight and creative discovery. This reframes memory consolidation: rather than merely defending against ...
226. Chebyshev Manifold Adaptation ​
Author: Jiawen Li
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17377v2 Announce Type: replace Abstract: The paper presents a new parameter-efficient adaptation method called ChebyMA (Chebyshev Manifold Adaptation). ChebyMA adopts weight matrices through a multi-surface superposition of Chebyshev polynomial bases evaluated on learnable coordinates and...
227. Decoder-Preserving Sparse Autoencoders: Which Readouts Survive Sparse Compression? ​
Author: Aniket Deshpande
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17425v2 Announce Type: replace Abstract: Sparse autoencoders (SAEs) compress model activations into sparse codes, but equal reconstruction error and sparsity can preserve different linearly decodable signals. We formalize this ambiguity as a matrix-valued distortion between optimal ridge-...
228. Enhancing Rubric-based RL via Self-Distillation ​
Author: Mingxuan Xia, Yuhang Yang, Chao Ye, Shuai Zhu, Shenzhi Yang, Guangcheng Zhu, Yuhang Zhang, Cheng Peng, Haobo Wang, Siqing Wang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18082v2 Announce Type: replace Abstract: Rubric-based RL has recently shown promise in improving LLMs on open-ended tasks. A widely recognized limitation of rubric-based RL is limited exploration: criteria that no rollout manages to satisfy (Unexplored Criteria, UC) receive no optimizatio...
229. When Are Scoring Rules Proper? Bridging Theory and Practice in Survival Model Evaluation ​
Author: John Zobolas, Raphael Sonabend, Riccardo De Bin, Johannes Piller, Philipp Kopper, Lukas Burk, Andreas Bender
Published: 7/22/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.AP, stat.TH
arXiv:2212.05260v4 Announce Type: replace-cross Abstract: Proper scoring rules encourage probabilistic predictions that match the true underlying distribution and are central to model evaluation, with increasing relevance in automated workflows such as AutoML. In survival analysis, however, their be...
230. Democratizing Advanced High-Throughput Imaging via Cross-Instrument Deep Learning-Enabled Modality Transfer ​
Author: Dominik Panek, Carina Rz\k{a}ca, Maksymilian Szczypior, Joanna Sorysz, Krzysztof Misztal, Zbigniew Baster, Zenon Rajfur
Published: 7/22/2026, 4:00:00 AM
Categories: eess.IV, cs.LG, q-bio.QM
arXiv:2403.18026v3 Announce Type: replace-cross Abstract: High-throughput imaging is often constrained by a trade-off between acquisition speed and image quality. Fast imaging modalities, such as wide-field fluorescence microscopy, enable large-scale data acquisition but suffer from reduced contrast...
231. Survival of the Cheapest: Cost-Aware Hardware Adaptation for Adversarial Robustness ​
Author: Charles Meyers, Mohammad Reza Saleh Sedghpour, Tommy L"ofstedt, Erik Elmroth
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CR, cs.CV, cs.LG, stat.AP
arXiv:2409.07609v3 Announce Type: replace-cross Abstract: Deploying adversarially robust machine learning systems requires continuous trade-offs between robustness, cost, and latency. We present an autonomic decision-support framework providing a quantitative foundation for adaptive hardware selecti...
232. Linear convergence of proximal descent schemes on the Wasserstein space ​
Author: Razvan-Andrei Lascu, Mateusz B. Majka, David \v{S}i\v{s}ka, {\L}ukasz Szpruch
Published: 7/22/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.PR
arXiv:2411.15067v2 Announce Type: replace-cross Abstract: We investigate proximal descent methods, inspired by the minimizing movement scheme introduced by Jordan, Kinderlehrer and Otto, for optimizing entropy-regularized functionals on the Wasserstein space. We establish linear convergence under fl...
233. Generalized Least Squares Kernelized Tensor Factorization ​
Author: Mengying Lei, Lijun Sun
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.CV, cs.LG
arXiv:2412.07041v4 Announce Type: replace-cross Abstract: Recovering incomplete multidimensional tensor-structured data is a fundamental task in many real-world applications. Smoothness-constrained low-rank tensor factorization effectively captures global and long-range correlations, but often strug...
234. Ab-initio simulation of excited-state potential energy surfaces with transferable deep quantum Monte Carlo ​
Author: Zeno Sch"atzle, P. Bern'at Szab'o, Alice Cuzzocrea, Mat\v{e}j Mezera, Frank No'e
Published: 7/22/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, physics.comp-ph
arXiv:2503.19847v2 Announce Type: replace-cross Abstract: The accurate quantum chemical calculation of excited states is a challenging task, often requiring computationally demanding methods. When entire ground and excited potential energy surfaces (PESs) are desired, for instance to predict the int...
235. A Geometry-Aware AI Emulator for the Coupled Whole Atmosphere from Earth Surface to the Ionosphere and Thermosphere ​
Author: Jiahui Hu, Wenjun Dong
Published: 7/22/2026, 4:00:00 AM
Categories: physics.space-ph, cs.LG
arXiv:2506.19340v4 Announce Type: replace-cross Abstract: Whole-atmosphere models such as WACCM-X resolve coupling from the Earth surface to the Mesosphere-Lower-Thermosphere (MLT), and Ionosphere-Thermosphere (IT) systems with expensive computational costs. Here we introduce CAM-NET, a geometry-awa...
236. Assistax: A Multi-Agent Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics ​
Author: Leonard Hinckeldey, Elliot Fosong, Rimvydas Rubavicius, Elle Miller, Trevor McInroe, Fan Zhang, Patricia Wollstadt, Stefano V. Albrecht, Subramanian Ramamoorthy
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA, cs.RO
arXiv:2507.21638v3 Announce Type: replace-cross Abstract: As embodied autonomous systems capable of assisting humans in daily activities remain a major goal for robotics, efficient and appropriate reinforcement learning (RL) simulation testbeds are increasingly important. Many common RL environments...
237. Node-as-Agent: Graph Agentic Network ​
Author: Minghao Guo, Xi Zhu, Qingyue Jiao, Xiujin Liu, Haochen Xue, Chong Zhang, Shuhang Lin, Jingyuan Huang, Ziyi Ye, Yongfeng Zhang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.MA
arXiv:2508.00429v5 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) have achieved remarkable success in graph-based learning by propagating information among neighbor nodes via predefined aggregation mechanisms. However, such fixed schemes often suffer from two key limitations. Fi...
238. Robust Belief-State Policy Learning for Quantum Network Routing Under Decoherence and Time-Varying Conditions ​
Author: Amirhossein Taherpour, Abbas Taherpour, Tamer Khattab, Mazen Hasna
Published: 7/22/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG, cs.NI
arXiv:2509.08654v2 Announce Type: replace-cross Abstract: Quantum network routing requires online decisions under probabilistic entanglement generation, finite quantum memories, decoherence, imperfect operations, and classical feedback, while the controller has incomplete knowledge of the physical s...
239. Low-Rank Evolutionary Deep Neural Networks via Adaptive Tangent-Space Reduction ​
Author: Jiahao Zhang, Shiheng Zhang, Guang Lin
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2509.16395v2 Announce Type: replace-cross Abstract: Evolutionary deep neural networks (EDNNs) solve time-dependent partial differential equations by evolving the neural-network parameters sequentially in time through a local least-squares problem. Their main computational bottleneck is that ea...
240. Pain in 3D: Generating Controllable Synthetic Faces for Automated Pain Assessment ​
Author: Xin Lei Lin, Soroush Mehraban, Abhishek Moturu, Babak Taati
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2509.16727v5 Announce Type: replace-cross Abstract: Automated pain assessment from facial expressions is crucial for non-communicative patient. Progress has been limited by two challenges: (i) existing datasets exhibit severe demographic and label imbalance due to ethical constraints, and (ii)...
241. Hyperdimensional Probe: Decoding LLM Representations via Vector Symbolic Architectures ​
Author: Marco Bronzini, Carlo Nicolini, Bruno Lepri, Jacopo Staiano, Andrea Passerini
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2509.25045v3 Announce Type: replace-cross Abstract: Despite their capabilities, Large Language Models (LLMs) remain opaque with limited understanding of their internal representations. Current interpretability methods either focus on input-oriented feature extraction, such as supervised probes...
242. Breaking the MoE LLM Trilemma: Dynamic Expert Clustering with Structured Compression ​
Author: Peijun Zhu, Ning Yang, Baoliang Tian, Jiayu Wei, Weihao Zhang, Haijun Zhang, Pin Lv
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.DC, cs.LG, cs.NE
arXiv:2510.02345v4 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) Large Language Models (LLMs) face a trilemma of load imbalance, parameter redundancy, and communication overhead. We introduce a unified framework based on dynamic expert clustering and structured compression to addre...
243. QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture ​
Author: Shvetank Prakash, Andrew Cheng, Mark Mazumder, Arya Tschand, Varun Gohil, Jeffrey Ma, Jason Yik, Zishen Wan, Jessica Quaye, Elisavet Lydia Alvanaki, Avinash Kumar, Chandrashis Mazumdar, Tuhin Khare, Alexander Ingare, Ikechukwu Uchendu, Radhika Ghosal, Abhishek Tyagi, Chenyu Wang, Andrea Mattia Garavagno, Sarah Gu, Alice Guo, Grace Hur, Luca P. Carloni, Tushar Krishna, Ankita Nayak, Amir Yazdanbakhsh, Vijay Janapa Reddi
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AR, cs.AI, cs.LG, cs.SE
arXiv:2510.22087v2 Announce Type: replace-cross Abstract: The field of computer architecture, which bridges high-level software abstractions and low-level hardware implementations, remains absent from current large language model (LLM) evaluations. To this end, we present QuArch (pronounced 'quark')...
244. DeepPAAC: A New Deep Galerkin Method for Principal-Agent Problems ​
Author: Michael Ludkovski, Changgen Xie, Zimu Zhu
Published: 7/22/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA
arXiv:2511.04309v3 Announce Type: replace-cross Abstract: We consider numerical resolution of principal-agent (PA) problems in continuous time. We formulate a generic PA model with continuous and lump payments and a multi-dimensional strategy of the agent. To tackle the resulting Hamilton-Jacobi-Bel...
245. Tumor-anchored deep feature random forests for out-of-distribution detection in lung cancer segmentation ​
Author: Aneesh Rangnekar, Harini Veeraraghavan
Published: 7/22/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2512.08216v4 Announce Type: replace-cross Abstract: Accurate segmentation of lung tumors from 3D computed tomography (CT) scans is essential for automated treatment planning and response assessment. Despite self-supervised pretraining on numerous datasets, state-of-the-art transformer backbone...
246. ImplicitRDP: An End-to-End Visual-Force Diffusion Policy with Structural Slow-Fast Learning ​
Author: Wendi Chen, Han Xue, Yi Wang, Fangyuan Zhou, Jun Lv, Yang Jin, Shirun Tang, Chuan Wen, Cewu Lu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2512.10946v2 Announce Type: replace-cross Abstract: Human-level contact-rich manipulation relies on the distinct roles of two key modalities: vision provides spatially rich but temporally slow global context, while force sensing captures rapid local contact dynamics. Integrating these signals ...
247. PRISP: Privacy-Safe Few-Shot Personalization via Lightweight Adaptation ​
Author: Junho Park, Dohoon Kim, Taesup Moon
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2601.06471v2 Announce Type: replace-cross Abstract: Large language model (LLM) personalization aims to adapt general-purpose models to individual users. Most existing methods, however, are developed under data-rich and resource-abundant settings, often incurring privacy risks. In contrast, rea...
248. RAPT: Model-Predictive Out-of-Distribution Detection and Failure Diagnosis for Sim-to-Real Humanoid Deployment ​
Author: Humphrey Munn, Brendan Tidd, Peter Bohm, Marcus Gallagher, David Howard
Published: 7/22/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2602.01515v2 Announce Type: replace-cross Abstract: Deploying learned control policies is risky because policies that appear robust in simulation can confidently enter out-of-distribution (OOD) states after Sim-to-Real transfer, causing silent failures and potential hardware damage. Existing a...
249. Beyond Content: Behavioral Policies Reveal Actors in Information Operations ​
Author: Philipp J. Schneider, Lanqin Yuan, Marian-Andrei Rizoiu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.SI, cs.LG
arXiv:2602.02838v2 Announce Type: replace-cross Abstract: The detection of online influence operations -- coordinated campaigns by malicious actors to spread narratives -- has traditionally depended on content analysis or network features. These approaches are increasingly brittle as generative mode...
250. Training and Simulation of Quadrupedal Robot in Adaptive Stair Climbing and Descending for Indoor Firefighting: An End-to-End Reinforcement Learning Approach ​
Author: Baixiao Huang, Baiyu Huang, Yu Hou
Published: 7/22/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2602.03087v2 Announce Type: replace-cross Abstract: Quadruped robots are used for primary searches during the early stages of indoor fires. A typical primary search involves quickly and thoroughly looking for victims under hazardous conditions and monitoring flammable materials. However, situa...
251. Segmented Continuous Optimization ​
Author: Teymur Aghayev
Published: 7/22/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2602.20857v2 Announce Type: replace-cross Abstract: Segmented curve fitting remains an essential approach for the comprehensive analysis of local patterns in non-stationary time-series data. However, traditional regression algorithms primarily focus on linear or polynomial functions, which can...
252. Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on Accelerators ​
Author: Zhengyang Su, Isay Katsman, Yueqi Wang, Ruining He, Lukasz Heldt, Raghunandan Keshavan, Shao-Chuan Wang, Xinyang Yi, Mingyan Gao, Onkar Dalal, Lichan Hong, Ed Chi, Ningren Han
Published: 7/22/2026, 4:00:00 AM
Categories: cs.IR, cs.CL, cs.LG
arXiv:2602.22647v2 Announce Type: replace-cross Abstract: Generative retrieval has emerged as a powerful paradigm for LLM-based recommendation. However, industrial recommender systems often benefit from restricting the output space to a constrained subset of items based on business logic (e.g. enfor...
253. Estimating near-verbatim extraction risk in language models with decoding-constrained beam search ​
Author: A. Feder Cooper, Mark A. Lemley, Christopher De Sa, Lea Duesterwald, Allison Casasola, Jamie Hayes, Katherine Lee, Daniel E. Ho, Percy Liang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2603.24917v2 Announce Type: replace-cross Abstract: Recent work shows that standard greedy-decoding extraction methods for quantifying memorization in LLMs miss how extraction risk varies across sequences. Probabilistic extraction -- computing the probability of generating a target suffix give...
254. Doctorina MedBench-ICD10: A Dialogue-Based Benchmark and Evaluation Framework for Agent-Based Medical AI ​
Author: Anna Kozlova, Stanislau Salavei, Pavel Satalkin, Hanna Plotnitskaya, Sergey Parfenyuk
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.MA
arXiv:2603.25821v2 Announce Type: replace-cross Abstract: We present Doctorina MedBench, a comprehensive evaluation framework for agent-based medical AI based on the simulation of realistic physician-patient interactions. Unlike traditional medical benchmarks that rely on solving standardized test q...
255. JD-BP: A Joint-Decision Generative Framework for Auto-Bidding and Pricing ​
Author: Linghui Meng, Chun Gan, Shengsheng Niu, Chengcheng Zhang, Chenchen Li, Chuan Yang, Yi Mao, Xin Zhu, Jie He, Zhangang Lin, Ching Law
Published: 7/22/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2604.05845v2 Announce Type: replace-cross Abstract: Auto-bidding services optimize real-time bidding strategies for advertisers under key performance indicator (KPI) constraints such as target return on investment and budget. However, uncertainties such as model prediction errors and feedback ...
256. Large Language Models Explore by Latent Distilling ​
Author: Yuanhao Zeng, Ao Lu, Lufei Li, Zheng Zhang, Yexin Li, Kan Ren
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2604.24927v2 Announce Type: replace-cross Abstract: Generating diverse responses is crucial for test-time scaling of large language models (LLMs), yet standard stochastic sampling mostly yields surface-level lexical variation, limiting semantic exploration. In this paper, we propose Explorator...
257. Lifting Embodied World Models for Planning and Control ​
Author: Alex N. Wang, Trevor Darrell, Pavel Izmailov, Yutong Bai, Amir Bar
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2604.26182v2 Announce Type: replace-cross Abstract: World models of embodied agents predict future observations conditioned on an action taken by the agent. For complex embodiments, action spaces are high-dimensional and difficult to specify: for example, precisely controlling a human agent re...
258. AgentJet: A Distributed Swarm Training Framework for Agentic Reinforcement Learning ​
Author: Qingxu Fu, Boyin Liu, Shuchang Tao, Zhaoyang Liu, Cheng Chen, Xuanfa Jin, Rong Zhu, Bolin Ding
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA
arXiv:2606.04484v2 Announce Type: replace-cross Abstract: Training reinforcement learning (RL) policies for large language model (LLM) agents requires optimizing multi-turn trajectories that interact with external environments. Existing training frameworks struggle with runtime failures, single-mode...
259. Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery ​
Author: Syed Rifat Raiyan, Mohsinul Kabir, Hasan Mahmud, Md Kamrul Hasan, Sophia Ananiadou
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CV, cs.LG
arXiv:2606.08728v3 Announce Type: replace-cross Abstract: Mathematical reasoning has long served as a stringent test of machine intelligence; over the past decade, it has moved from a niche problem within NLP to one of the most consequential AI frontiers. This survey provides a unified account of th...
260. The Correctness Illusion in LLM-Generated GPU Kernels ​
Author: Dipankar Sarkar
Published: 7/22/2026, 4:00:00 AM
Categories: cs.SE, cs.DC, cs.LG
arXiv:2606.20128v2 Announce Type: replace-cross Abstract: Benchmarks for LLM-generated GPU kernels (KernelBench, TritonBench, GEAK) score correctness through fixed-shape, small-sample allclose-style checks. The number of inputs varies between benchmarks. The shape, dtype, and tolerance are fixed for...
261. Signed Evidence Flow: Conflict-Aware and Stability-Calibrated Data Analysis ​
Author: Jeffery Opoku, David Banahene
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2606.21875v2 Announce Type: replace-cross Abstract: Modern data analysis usually gives a prediction without showing whether the evidence behind it is clear, conflicting, or stable. Two cases can have the same fitted confidence even when one has mostly agreeing evidence and the other has strong...
262. Flow-Corrected Thompson Sampling for Non-Stationary Contextual Bandits ​
Author: Ali Baheri
Published: 7/22/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2606.23933v2 Announce Type: replace-cross Abstract: We study non-stationary linear contextual bandits where the reward model drifts over time, rendering classical contextual bandit algorithms brittle because historical data becomes systematically biased. We propose Flow-Corrected Thompson Samp...
263. Reclaim Evaluation: A Lossy Memory Is Worse Than an Empty One ​
Author: Alex Kwon
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2606.25449v5 Announce Type: replace-cross Abstract: A language model's memory can be worse than no memory at all when the model or its interface is disposed to act on it: a memory that keeps a wrong conclusion but drops the work behind it leads a model to re-emit the stale value as a confident...
264. EvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures ​
Author: Bu\u{g}ra Alperen Ulu{\i}rmak, Rifat Kurban
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.SE
arXiv:2606.30219v3 Announce Type: replace-cross Abstract: LLM evaluation and AI safety face a shared measurement problem: benchmark scores, reward-model signals, and reported safety metrics can improve while the latent properties they are meant to represent remain difficult to verify. This paper com...
265. Forensic Trajectory Signatures for Agent Memory Poisoning Detection ​
Author: Jun Wen Leong
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2606.30566v2 Announce Type: replace-cross Abstract: We discover a behavioral invariant in LLM agents under persistent memory poisoning and characterize its deployment boundary. In architectures where retrieval is routed through observable memory-tool invocations, successful attacks require cal...
266. Computer vision-based neural networks for radioisotope identification in urban environments ​
Author: Masen Bachleda, Alea Minar, Ayush Panigrahy, Peter Lalor
Published: 7/22/2026, 4:00:00 AM
Categories: physics.ins-det, cs.LG
arXiv:2607.00270v2 Announce Type: replace-cross Abstract: Algorithm development for radioisotope identification in mobile urban search scenarios face significant challenges from non-uniform backgrounds, momentary source encounters, and severe class imbalance between rare threat signatures and backgr...
267. Vidu S1: A Real-Time Interactive Video Generation Model ​
Author: Jintao Zhang, Kai Jiang, Jintao Chen, Xu Wang, Yang Luo, Yuji Wang, Dechuang Chen, Jungang Li, Chengyang Ye, Marco Chen, Hongzhou Zhu, Min Zhao, Yuxuan Jiang, Zhengkun Huang, Chendong Xiang, Kaiwen Zheng, Haoxu Wang, Xiaohang Wang, Qi Jia, Xin Chen, Yimin Chen, Youhe Jiang, Fangcheng Fu, Zhijie Deng, Fan Bao, Jianfei Chen, Jun Zhu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.03118v2 Announce Type: replace-cross Abstract: We introduce Vidu S1, a real-time interactive video generation model supporting voice control of digital characters. Users can control video generation content at any moment through voice instructions. Vidu S1 supports infinite-length real-ti...
268. Don't Blame the Large Language Model: How Agent Harness Evolution Shapes Coding Agent Quality ​
Author: Oussama Ben Sghaier, Hao Li, Bram Adams, Ahmed E. Hassan
Published: 7/22/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2607.03691v2 Announce Type: replace-cross Abstract: Coding agents, autonomous systems that use large language models (LLMs) to resolve software engineering tasks, rely on agent harness: a middleware layer in between a developer and a large language model that orchestrates system prompts, tool ...
269. Context-Masked Truncated Reasoning Audits for Answer-Key Dependence in LLM Tutors ​
Author: Bonan Shen, Dingyan Shang, Youting Wang, Tao Ning, Bowen Liu
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.04572v2 Announce Type: replace-cross Abstract: Large language model (LLM) tutors may have access to teacher notes, answer keys, rubrics, or retrieved solutions while producing student-facing explanations. We study whether truncated reasoning probes can distinguish direct access to such pr...
270. LP-SFT: Local-Preserving Supervised Fine-Tuning via Multimodal Entropy Structure ​
Author: Yueyang Wang, Baolong Bi, Shuo Lu, Jingyuan Zhang, Jiajun Shi
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.04733v2 Announce Type: replace-cross Abstract: Supervised fine-tuning (SFT) is the standard approach for adapting pretrained language models to downstream domains, yet it often improves target-domain behavior at the cost of degrading pre-existing capabilities. Standard cross-entropy fine-...
271. Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages ​
Author: Lucas Pinto
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.06596v2 Announce Type: replace-cross Abstract: Trusted monitoring is a central defense in AI control: a cheaper trusted model scores an untrusted model's actions for sabotage, and the most suspicious are audited or deferred. Such monitors are evaluated against one or two untrusted models,...
272. Finding a stationary point of a stochastic convex problem ​
Author: Felipe Areces, John Duchi, Malo Sommers
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC
arXiv:2607.06883v2 Announce Type: replace-cross Abstract: We consider the problem of finding stationary points for stochastic convex optimization problems. Rather than surrogates to stationarity, such as a proximity-to-stationarity guarantee or small gradient of the Moreau envelope, we ask for a str...
273. One mechanism for many mental spaces: a shared router over a value slot in language models ​
Author: Oliver Steele, Jiangtao Wen, Yuxing Han
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.10248v2 Announce Type: replace-cross Abstract: Language builds discourse contexts other than the actual: a painting, a belief, a memory, a hypothetical. Each is a mental space in which the same entity can take a different value, as when a flower is red in reality but purple in a portrait....
274. Belief-reality separation lives in routing over a shared value slot in language models ​
Author: Oliver Steele, Jiangtao Wen, Yuxing Han
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.11945v2 Announce Type: replace-cross Abstract: Capable language models hold what a character believes apart from what is true: told "Anna believes the cup is blue; in reality it is red," they answer blue about Anna and red about the world. Where in the computation does that separation liv...
275. Deep-learning Causal Retrieval Optimization for Efficient e-commerce Distribution in Pinterest ​
Author: Junpeng Hou, XianXing Zhang, Sai Xiao, Derek Cheng, Darren Reger, Olafur Gudmundsson, Mehdi Ben Ayed, Zhiqing Rao, Huizhong Duan
Published: 7/22/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.14161v2 Announce Type: replace-cross Abstract: Pinterest is where people turn inspiration into action as users browse ideas, then take steps toward realization, often by discovering shoppable content. To support this journey, we must distribute commerce content when it helps, not when it ...
276. NexForge: Scaling Agent Capabilities through Requirement-Driven Task Synthesis for LLMs ​
Author: Jiarong Zhao, Zhikai Lei, Zhiheng Xi, Rui Zheng, Hang Yan, Jie Zhou, Qin Chen, Liang He
Published: 7/22/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2607.14186v4 Announce Type: replace-cross Abstract: Scaling executable agent training data for LLM post-training is bottlenecked by substrate-bound methods that tie task generation to predefined tools, repositories, or skill graphs: expanding coverage requires manual substrate engineering, eac...
277. RetroAgent: Harnessing LLMs to Search Over Structured Memory for Agentic Retrosynthesis Planning ​
Author: Yanqiao Zhu, Jingru Gan, Xiaoqi Sun, Fang Sun, Yidan Shi, Md Mofijul Islam, Chao Shang, Wenhao Gao, Connor W. Coley, Yizhou Sun, Wei Wang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.14512v2 Announce Type: replace-cross Abstract: Multi-step retrosynthesis planning seeks to decompose a target molecule into commercially available building blocks through a sequence of feasible reactions. The vast combinatorial search space makes this task challenging even for expert chem...
278. MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators ​
Author: Yushi Huang, Xiangxin Zhou, Jun Zhang, Liefeng Bo, Tianyu Pang
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.15273v2 Announce Type: replace-cross Abstract: MeanFlow generators achieve fast few-step sampling by predicting average velocities over time intervals, making them attractive for efficient generation. Reinforcement learning (RL) has become a powerful way to align diffusion and flow models...
279. It Depends on the Dataset: When a Brain-Encoding Model's Predicted Responses Beat Their Visual Backbone for Video Memorability ​
Author: Carson Rodrigues
Published: 7/22/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.16292v2 Announce Type: replace-cross Abstract: Brain-encoding foundation models predict fMRI responses to video, audio, and text well enough to win the Algonauts 2025 challenge. We ask whether their predicted responses, obtained with no scanner, are a useful feature lens for a downstream ...
280. Lipschitz Continuity in Deep Learning: A Systematic Review of Theoretical Foundations, Estimation Methods, Regularization Approaches, and Certifiable Robustness ​
Author: R'ois'in Luo, James McDermott, Colm O'Riordan
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.16329v2 Announce Type: replace-cross Abstract: Lipschitz continuity is a fundamental property of neural networks that characterizes their sensitivity to input perturbations. It plays a pivotal role in deep learning, governing \textbf{robustness}, \textbf{generalization} and \textbf{optimi...
281. AEVAL: From Anecdotal to Deterministic Testing for Agentic Skill Workflows ​
Author: Tejas Singh Anand, Yuet Ying Christina Wang, Wanting Jiang, Steve Masson, Tian Zheng, Bingjie Zhou
Published: 7/22/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG, cs.PF
arXiv:2607.16345v2 Announce Type: replace-cross Abstract: Modern agentic systems increasingly rely on skills: installable packages of natural language and code that teach an LLM agent to perform a domain task. As skill repositories grow, developers need automated quality signals on every change, yet...
282. Quantifying Diversity of Thought: A Predictive Law of Weighted LLM Ensemble Lift ​
Author: Junade Ali
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO, cs.MA
arXiv:2607.17384v2 Announce Type: replace-cross Abstract: This paper provides an experimentally verified formal law for calculating the uplift that diversity of thought provides in Large Language Model (LLM) ensembles. From first principles, we derive an exact decomposition of LLM ensemble lift into...
283. Kernel Regression with Tensor Trains and Hadamard Overparameterization ​
Author: Duc Thien Nguyen, Konstantinos Slavakis, Eleftherios Kofidis, Dimitris Pados
Published: 7/22/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, eess.SP
arXiv:2607.17390v2 Announce Type: replace-cross Abstract: Kernel regression with tensor trains and Hadamard overparameterization (KReTTaH) is introduced as a training-data-free, interpretable, and nonparametric framework for multi-way data imputation. The imputation problem is reformulated as regres...
284. Can AI Agents Really Complete RTL-to-GDS? Lessons from Benchmarking Tool-Interactive EDA Workflows ​
Author: Jinyuan Deng, Zhengrui Chen, Xufeng Wei, Tianyu Xing, Chenyi Wen, Cheng Zhuo
Published: 7/22/2026, 4:00:00 AM
Categories: cs.AI, cs.AR, cs.LG
arXiv:2607.17528v2 Announce Type: replace-cross Abstract: LLM-driven agent systems have emerged as a promising paradigm for electronic design automation (EDA), demonstrating strong potential for automating complex design workflows. However, existing evaluations primarily examine individual language ...