Skip to content

arXiv cs.LG - 2026-07-23 ​

268 items collected.


1. Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions ​

Author: David R. Wessels, Farhad Ramezanghorbani, David W. Romero, Alireza Moradzadeh, Olivia Viessmann, Maksim Zhdanov, John St. John, Ken Janik, David M Knigge, Yucheng Tang, Erik J Bekkers, Saee Gopal Paliwal
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, stat.ML

arXiv:2607.19378v1 Announce Type: new Abstract: Subquadratic alternatives to attention require compromises when applied to multi-dimensional data: standard convolutions lack global receptive fields and input dependency, while recurrent models require rasterizing data such as images, volumes, and par...

📖 Read original article


2. Bayesian Wind Tunnels for Model Selection ​

Author: Siddhartha R Dalal, Vishal Misra, Abhay Parekh
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.19379v1 Announce Type: new Abstract: Prior work has shown that transformers can perform exact Bayesian filtering within a fixed hypothesis class. Can they also perform Bayesian model selection -- identifying the correct hypothesis class from data? We introduce model-selection Bayesian win...

📖 Read original article


3. CruiseBench: A Real-Flight-Aligned N-CMAPSS Benchmark for Engine RUL Prediction ​

Author: Pu Cheng, Qiang Miao
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19380v1 Announce Type: new Abstract: Remaining useful life (RUL) prediction estimates how long an engine can continue safe operation and is central to maintenance planning. N-CMAPSS extends C-MAPSS by simulating run-to-failure aero-engine trajectories using recorded real-flight profiles a...

📖 Read original article


4. Air Quality Arena: A Large-Scale Multi-Region Ground Monitoring Dataset and Benchmark for Air Quality Forecasting with Time-Series Foundation Models ​

Author: Rishi Bharadwaj, Manik Gupta, Pandarasamy Arjunan
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19381v1 Announce Type: new Abstract: Air pollution causes an estimated 7.9 million premature deaths annually, making accurate forecasting a critical public health priority. Machine learning is increasingly being applied to forecast air pollution levels, yet existing benchmarks remain narr...

📖 Read original article


5. Challenges of Explainability in Continual Learning for Time Series Forecasting ​

Author: Quentin Besnard (RFAI), Emmanuel Doumard (BDTLN), Nicolas Labroche (LIFAT, BDTLN), Nicolas Ragot (RFAI), Nicolas Ringuet (BDTLN)
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19382v1 Announce Type: new Abstract: Deep learning models have shown strong potential for time series forecasting, yet their deployment in real-world environmental monitoring remains challenging due to non-stationary dynamics and limited explainability. In this work, we investigate explai...

📖 Read original article


6. SUM: Unified Geometric Surgery on Spatio-Temporal Adaptation Vectors for Federated Class Incremental Learning ​

Author: Jaeik Kim, Jaeyoung Do
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19384v1 Announce Type: new Abstract: Real-world intelligent systems often require both distributed collaboration across data-isolated clients and continual adaptation to evolving tasks. This setting naturally gives rise to Federated Class Incremental Learning (FCIL), which combines Federa...

📖 Read original article


7. STN-TGAT: Top-K Portfolio Construction via Prior-Guided Graph Attention with Learnable Soft-Threshold Sparsification ​

Author: Haoran Guo, Yutong Lu, Li Zhang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19385v1 Announce Type: new Abstract: This paper tackles the problem of stock ranking and portfolio construction under realistic investment settings by jointly modeling temporal dynamics and cross-sectional dependencies. We propose the Soft-Threshold NMI-prior Transformer Graph Attention N...

📖 Read original article


8. Building Fast, Evaluating Slow: Pipeline Choices Dominate Autointerpretability Score Variance ​

Author: Sinie van der Ben, Neele Roch, Anna Hedstr"om, Mennatallah El-Assady
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.19386v1 Announce Type: new Abstract: Cross-paper comparison of sparse autoencoder (SAE) interpretability often relies on autointerpretability scores. In this evaluation pipeline, a language model (LM) explains each feature, and another LM scores the explanation. For these comparisons to b...

📖 Read original article


9. Scale-Aware Learning of Chaotic Dynamics on Unstructured Meshes via Binned Spectral Losses ​

Author: Kanad Sen, Romit Maulik
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, physics.comp-ph, physics.flu-dyn

arXiv:2607.19387v1 Announce Type: new Abstract: Surrogate modeling for high-dimensional nonlinear dynamical systems that exhibit chaos requires mechanisms that preserve not only pointwise accuracy but also the scale-dependent structure of physical fields. Bandwise spectral power losses, such as the ...

📖 Read original article


10. Neural Operator Surrogates for Two-Dimensional Neutron Flux Estimation ​

Author: Japan K. Patel, Barry D. Ganapol, Anthony Magliari, Matthew C. Schmidt, Todd A. Wareing
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19388v2 Announce Type: new Abstract: This work extends our one-dimensional single-sweep neural-operator studies to two dimensions. We consider one-group transport with isotropic scattering. As in the one-dimensional work, we use Fourier neural operators (FNOs) to approximate the high-fide...

📖 Read original article


11. The Orthogonalized Read Is a Removable Training Scaffold for Recurrent Memory ​

Author: Keston Aquino-Michaels
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19390v2 Announce Type: new Abstract: A recent report finds that orthogonalizing the mLSTM memory matrix at read time (five Newton-Schulz iterations, trained through) substantially improves noisy associative recall. The effect replicates, but it is not a memory improvement. Training on thi...

📖 Read original article


12. LAARA: Layer-Aware Adaptive Rank Allocation for Parameter-Efficient Fine-Tuning ​

Author: Ashutosh Tripathi, Surya Deep Singh, Pranab Sahoo, Sriparna Saha
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19391v1 Announce Type: new Abstract: Low-Rank Adaptation is widely used for parameter-efficient fine-tuning, yet existing methods typically assign the same adapter rank to every transformer layer despite their heterogeneous adaptation requirements. In this work, we show theoretically and ...

📖 Read original article


13. Predicting Groundwater Arsenic Concentrations Using Graph Neural Networks ​

Author: William Xing, Stephanie Yang, Aarush Bandemegal, Anushree Misra, Ananya Kalapatapu, Brennan Lagasse, Kevin Zhu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19392v1 Announce Type: new Abstract: Arsenic contamination in groundwater presents a longstanding public health crisis in the United States, especially for households depending on private wells. Accurate and spatially informed prediction of arsenic concentration is vital to identify high-...

📖 Read original article


14. Decodable but Not Detectable: A Leakage Fingerprint for Near-OOD Benchmarks ​

Author: Vishnu Bindu Balachandran
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CV

arXiv:2607.19393v1 Announce Type: new Abstract: While auditing a perturbation-based OOD detector on a document benchmark, we recorded an AUROC of 0.326 -- well below the 0.5 chance level. The cause is a benchmark leak: the designated "OOD" class is one the model was trained on, so its examples sit i...

📖 Read original article


15. Cross-Subject Semantic Decoding with Shared-Space Alignment for Generalized Neural Representation Learning ​

Author: Ji-Hoon Heo, Aleksandra Joanna Wisniewska, Seo-Hyun Lee, Seong-Whan Lee
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19394v1 Announce Type: new Abstract: Generalizing across subjects remains challenging in invasive neural recordings because electrode configurations, anatomical structures, and neural signal patterns vary substantially across individuals. To investigate such inter-subject variability, we ...

📖 Read original article


16. From Trajectories to Prefixes: Reusing Teacher Trajectories via Replayed Prefixes and Online Continuation ​

Author: Yihan Wang, Zhong Guan, Haoran Sun, Jiale Huang, Likang Wu, Hongke Zhao
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19395v1 Announce Type: new Abstract: Small language models are attractive backbones for interactive agents, but direct distillation from strong teacher trajectories often turns rich multi-turn behavior into one-shot imitation targets. This is inefficient in long-horizon environments, wher...

📖 Read original article


17. Memory Merge DQN: Sensitivity Weighted Target Updates for Stable Value Learning ​

Author: Adrian Ly, Richard Dazeley, Peter Vamplew, Sunil Aryal, Francisco Cruz
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19397v1 Announce Type: new Abstract: Deep Q-networks use target networks to stabilise bootstrapped value learning, but the standard hard copy update also introduces a tradeoff. Holding the target network fixed, improves short term stability, yet each hard update abruptly replaces the targ...

📖 Read original article


18. Leveraging Offline Supervision for Efficient and Generalizable Reinforcement Learning in Large-Scale Vision-Language-Action Models ​

Author: Dmitriy Poyarkov, Aleksei Staroverov, Aleksandr I. Panov
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.19399v1 Announce Type: new Abstract: It is commonly observed that online reinforcement learning (RL) produces better-performing strategies than offline methods across a broad range of performance measures. In particular, RL-trained policies exhibit stronger out-of-distribution (OOD) behav...

📖 Read original article


19. Predictive single cell foundation model for gene regulation and aging with privacy-preserving tabular learning ​

Author: Jiayuan Ding, Jianhui Lin, Ziyang Miao, Nils Mechtel, Shiyu Jiang, Yixin Wang, Zhaoyu Fang, Jorge D. Martin-Rufino, Chen Weng, Reuben Saunders, Weize Xu, Jonathan S. Weissman, Min Li, Jiliang Tang, Wei Ouyang, Yuancheng Ryan Lu, Xiaojie Qiu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, q-bio.GN

arXiv:2607.19400v1 Announce Type: new Abstract: Pre-trained foundation models (FMs) have begun transforming single-cell genomics, but scaling them raises privacy concerns. Moreover, unlike text data, single-cell data is unordered and exhibits a unique tabular structure that current single-cell FMs o...

📖 Read original article


20. When Does Consensus Beat Voting? A Critical Analysis of Statistical Label Fusion in Medical Image Segmentation ​

Author: Renjie He
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, physics.med-ph

arXiv:2607.19402v1 Announce Type: new Abstract: This paper provides a rigorous, self-contained investigation of consensus segmentation. We derive the mathematical foundations from first principles -- the generative model, EM algorithm, Van Leemput's marginalization analysis, identifiability conditio...

📖 Read original article


21. Recovering Clinical Utility Under Differential Privacy: Empirical Validation of Adaptive Federated Aggregation on Heterogeneous Cardiovascular Datasets ​

Author: Rodrigo Tertulino, Laercio Alencar, Ricardo Almeida
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR, cs.CY

arXiv:2607.19403v1 Announce Type: new Abstract: Validating federated learning frameworks on real clinical data is an essential step between proof-of-concept demonstrations in controlled synthetic environments and deployment in real multicenter healthcare settings. A prior architectural study by the ...

📖 Read original article


22. Structured Latent Space Modeling over Multi-Scale Temporal Patches for Multivariate Time Series Forecasting ​

Author: Xingsheng Chen, Deyu Yi, Siu-Ming Yiu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19404v1 Announce Type: new Abstract: Multivariate time series encode structural patterns that unfold across multiple temporal scales, yet most forecasting backbones treat learned representations as transient byproducts of prediction, leaving the organizational geometry of these patterns u...

📖 Read original article


23. Reproducing Recurrent Transformers: The CoTFormer ​

Author: Aras Kavuncu, Bryan Vullo, Alberto Berni
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19405v1 Announce Type: new Abstract: The CoTFormer architecture formalizes Chain-of-Thought as a form of recurrent latent computation, preserving intermediate states as attendable representations to mimic explicit reasoning traces. In this work, we evaluate CoTFormer and its structural va...

📖 Read original article


24. NMR Elucidation as an Agentic Search Problem, Not a Modeling Problem ​

Author: Irina Espejo Morales, Damon Hinz, Marvin Alberts, Geraud Krawezik, Haewon Jeong, Shirley Ho
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19406v1 Announce Type: new Abstract: Structural elucidation from Nuclear Magnetic Resonance (NMR) data remains a fundamental bottleneck across chemistry, materials science, and biology. We demonstrate that an agentic AI system can perform this task at a level comparable to graduate-level ...

📖 Read original article


25. Reward-Aware Population Scaling of Evolutionary Strategies in LLM Fine-Tuning ​

Author: Sung Cho, Gyubin Han
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19408v1 Announce Type: new Abstract: Using Evolutionary Strategies (ES) for fine-tuning large language models is attractive because it is memory-efficient, parallel, and compatible with black-box or discrete rewards. Yet its population-size conclusions conflict sharply: fine-tuning with c...

📖 Read original article


26. Adaptive Multi-Expert Graph Transformer for Interpretable EEG-Based Diagnostics ​

Author: Maryam Rahimimovassagh, Md Elias Hossain, Ivan Garibay, Niloofar Yousefi
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19429v1 Announce Type: new Abstract: Electroencephalographic (EEG) abnormalities arise from dynamic changes in neural synchrony across spatial and temporal scales, yet many computational approaches reduce these dynamics to static features. We present a Spatial Multi-Expert Graph Transform...

📖 Read original article


27. Unlearning as Distribution Restoration: A Controlled Counterfactual Study, a Validated Selective Screen, and the Limits of Oracle-Free Certification ​

Author: Sen Yang, Yuen-Hei Yeung
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19442v1 Announce Type: new Abstract: Machine unlearning is commonly evaluated by matching a retrained oracle on trained probes. In a controlled nonce-fact testbed with a matched retraining reference, we find this criterion can favor methods that retain held-out knowledge: candidates it ra...

📖 Read original article


28. Marine Engine Fault Dataset: Open-Access Data under Controlled Reference and Fault Scenario Conditions ​

Author: Ahmad BahooToroody, Oleksiy Bondarenko, Mohammad Mahdi Abaei, Niki Yoichi, Enrico Zio
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SP, eess.SY

arXiv:2607.19444v1 Announce Type: new Abstract: Open-access datasets for marine-engine predictive maintenance remain scarce, particularly those from controlled fault experiments with documented operating conditions, subsystem-level interventions and system-level measurements. This work presents the ...

📖 Read original article


29. Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents ​

Author: Aarushi Singh
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2607.19449v1 Announce Type: new Abstract: Evaluation frameworks for tool-augmented LLM agents focus overwhelmingly on capability metrics or explicit tool crashes, leaving silent infrastructure failures and HTTP 200 responses with empty, null, or malformed payloads largely unaudited. We introdu...

📖 Read original article


30. REGEN: Replay-recycling for Expert-to-Generalist distillation with Offline Reinforcement Learning ​

Author: Yunjie Chen, Xiaoxin Chen, Fang Wang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19450v1 Announce Type: new Abstract: Large-scale online reinforcement learning (RL) is the predominant means of eliciting advanced abilities including long-term reasoning and agentic tool use in large language models (LLMs). However, continuing to scale it across vast task domains of inte...

📖 Read original article


31. Predictive Extrema, Unprofitable Policies: An AI-Assisted Audit of Candle-Based Binance Spot Timing Models ​

Author: Ayoub Jadouli
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-fin.ST, q-fin.TR

arXiv:2607.19453v1 Announce Type: new Abstract: We audit whether candle-based machine-learning models can turn predictions of cryptocurrency extrema or short-horizon outcomes into positive Binance Spot paper policies after assumed costs. Numerical results come from scripted fixed-seed model runs and...

📖 Read original article


32. Generating Bearing Vibration Signals at User-Specified Fault Probabilities Using PR-GAN and Counterfactual Methods ​

Author: Seyed Mohammadreza Alavi, Ardeshir Shojaeinasab, Reza Jalayer, Masoud Jalayer, Behnam Bahrak
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19455v1 Announce Type: new Abstract: In bearing vibration datasets, most samples receive predicted fault probabilities close to 0 or 1, while samples with intermediate (gray-zone) probabilities are rare. Such borderline samples are important because they reflect conditions in which mainte...

📖 Read original article


33. MoA-Structured Decode Attention DNF Derivation, KV-Cache Accumulation, GQA/MQA, and OpenACC Kernel ​

Author: Lenore Mulin, Gaetan Hains
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19456v1 Announce Type: new Abstract: We derive four memory-optimal inference artifacts for transformer attention using the Mathematics of Arrays (MoA), each following directly from the forward-pass Denotational Normal Form (DNF) of with the query-row index fixed to the current decode step...

📖 Read original article


34. Total Variation Distance Estimation in Autoregressive Models ​

Author: Eric Price, Kevin Tian, Zhiyang Xun, Yusong Zhu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, stat.ME, stat.ML

arXiv:2607.19510v1 Announce Type: new Abstract: Modern LLM deployments use a number of implementation choices and inference optimizations (e.g., batching, custom kernels, and quantization) on top of fixed weights, so two engines serving "the same model" can produce meaningfully different distributio...

📖 Read original article


35. Do Sheaf Neural Networks Use Holonomy? A Measure--Intervene--Control Study ​

Author: Ankit Grover, R'emi Bourgerie
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19514v1 Announce Type: new Abstract: Geometric architectures are often justified by internal mechanisms such as rotations, yet task performance alone cannot show whether those mechanisms drive predictions. Using sheaf neural networks (SNNs) as a testbed, we introduce the first basis-indep...

📖 Read original article


36. Geospatial Diffusion-based Evolution Synthesis (GeoDES) for Storm-Centered Weather Augmentation ​

Author: Sonia Cromp, Satya Sai Srinath Namburi GNVV, Youran Wang, Grace Kisslinger, Frederic Sala, James Booth, Allegra LeGrande
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.19522v1 Announce Type: new Abstract: While machine learning-based weather models hold significant promise, they struggle to predict the detailed structure of large-scale weather systems such as cyclonic storms. Regional models are constrained by limited historical records within fixed geo...

📖 Read original article


37. SynPre-FL: Synthetic data-driven pretraining integrated Federated Learning training framework ​

Author: Akarsh K Nair, Muhammad Arifur Rahman, Nicholas Shopland, Andy Burton, Jun He, Yuan Shen, David Baldwin, Emma O'Dowd, Amna Burzic, Mufti Mahmud, David J. Brown
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC

arXiv:2607.19524v1 Announce Type: new Abstract: Federated learning (FL) offers a promising approach to privacy-preserving clinical risk prediction, but its deployment remains limited by restricted data sharing, client heterogeneity, class imbalance, and the lack of realistic tabular electronic healt...

📖 Read original article


38. The C-index illusion: discrimination without calibration in published survival models ​

Author: Rafael da Silva, Danilo Alvares
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19526v2 Announce Type: new Abstract: Recent work has argued normatively, on synthetic data, that evaluating survival models by discrimination alone (concordance index) yields systematically misleading model comparisons, because the metric ignores calibration and time-dependent accuracy. W...

📖 Read original article


39. Trustworthy Privacy-Preserving Multimodal Federated Learning for Personalised Breast Cancer Prediction ​

Author: Ruth Amey, Muhammad Arifur Rahman, Taha Osman, Nicholas Shopland, Andy Burton, Mufti Mahmud, David J. Brown
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19532v1 Announce Type: new Abstract: Federated learning has emerged as a potential solution to privacy concerns associated with using sensitive health data for training predictive models, particularly in personalised cancer care. This research investigates whether federated learning can s...

📖 Read original article


40. Agent-Centric Animal Pose Forecasting ​

Author: Eyrun Eyjolfsdottir, Kristin Branson
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19548v1 Announce Type: new Abstract: Understanding animal behavior at an algorithmic level -- what animals attend to, how they form internal models and plans, and how this maps to action -- remains a central challenge in neuroscience and ethology. Data-driven generative models offer a pat...

📖 Read original article


41. End-to-End Differential Privacy in Training Deep Neural Network Classifiers ​

Author: Huaiyuan Rao, Calvin Hawkins, Alexander Benvenuti, Matthew Hale
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2607.19580v1 Announce Type: new Abstract: Differentially private machine learning enables model training on sensitive data while ensuring that individual data is unlikely to be recoverable from the parameters of the resulting model. However, existing work often privatizes both training inputs ...

📖 Read original article


42. The Mechanism Matters: When Knowledge Graphs Help Reinforcement Learning ​

Author: Mohammed Sameer Syed
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19616v1 Announce Type: new Abstract: Knowledge graphs (KGs) are widely used to inject prior knowledge into reinforcement learning (RL), yet the literature is dominated by single-domain, positive-result method papers, so we lack a systematic account of when KG structure helps an agent, whe...

📖 Read original article


43. SCPP: A Unified Python Library for Soft Clustering ​

Author: Kiyan Rezaee, Morteza Ziabakhsh, Artin Bahrampour, Seyed Mohammad Ghoreishi, Asal Khaje, Ali Sajedifar, Manny Chalak, Ava Zerafatangiz, Sadegh Eskandari
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19620v1 Announce Type: new Abstract: In this paper, we present SCPP (Soft Clustering Python Package), an open-source Python framework for soft clustering. SCPP establishes a canonical, scikit-learn-compatible estimator interface that standardizes model training, prediction, membership rep...

📖 Read original article


44. HypEMBER: Hypernetwork-based Ensemble for Robust Policy Learning of Parametrized Dynamical Systems ​

Author: Nicol`o Botteghi, Gabriele Pascali, Urban Fasel, Andrea Manzoni
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19628v1 Announce Type: new Abstract: In this work we investigate reinforcement learning (RL) as a framework for the robust control of parametrized dynamical systems in presence of measurements and model uncertainties. High-dimensional state spaces, expensive numerical solvers, the partial...

📖 Read original article


45. Anatomy of a Sound Neural Reasoner: One-Shot Amortization, First-Pass Poisoning, and Search Inertness in Clue-Rich Completion ​

Author: Aleksey Komissarov
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19635v1 Announce Type: new Abstract: Neural solvers are built to deduce, branch, and revise intermediate states. The Lattice Deduction Transformer (LDT) appears to do exactly that. In clue-rich Sudoku, it does not: one forward pass commits essentially the entire grid (every blank cell on ...

📖 Read original article


46. Expert-Guided Forecast Editing for Time-Series Foundation Models ​

Author: Hung Le, Minh Hoang Nguyen, Manh Nguyen, Huu Hiep Nguyen, Dai Do
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19659v1 Announce Type: new Abstract: Time-series foundation models can forecast across heterogeneous domains without task-specific training, but their forecasts are fixed once produced and cannot directly incorporate task-specific expert feedback. We study expert-guided forecast editing: ...

📖 Read original article


47. Efficient Clustering with Provable Guardrails for LLM Inference at Scale ​

Author: Longshaokan Wang, Wai Tsang Keung, Punit Ghodasara, Roman Wang, Ali Dashti, Francesc Moreno-Noguer
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.19704v1 Announce Type: new Abstract: Scaling LLM-based applications to millions of users is bottlenecked by the inference cost and latency of modern foundation models. A natural fix is to cluster the inputs and call the LLM only on cluster representatives, letting other members inherit th...

📖 Read original article


48. How Fast Can Reward Models Score? A Systems Study of C++ and PyTorch Inference Runtimes for RLHF ​

Author: Venkata Naga Sai Vishnu Rohit Pulipaka, Anish Katta, Deva Rohit Reddy Peddireddy
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19712v1 Announce Type: new Abstract: In RLHF pipelines, reward scoring blocks policy updates. Slow scoring bottlenecks the entire loop, since no update runs until every rollout gets a score. And yet most setups just default to PyTorch eager mode or torch.compile, no one checks if that's a...

📖 Read original article


49. Koopman Dreamer: Spectrally Constrained Latent Dynamics for Stable World-Model Imagination ​

Author: Jiaqi Li, Xinglong Zhang, Haibin Xie, Yixing Lan, Wei Pan, Xin Xu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2607.19719v1 Announce Type: new Abstract: Latent world models improve sample efficiency in continuous control by optimizing policies over imagined latent trajectories, but common neural transitions offer limited direct control over modal persistence and error accumulation in long rollouts. We ...

📖 Read original article


50. Analytic Distribution of Classifier-Free Guidance for Schedule Design ​

Author: Enze Jiang, Zheng Ma
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.19725v1 Announce Type: new Abstract: Classifier-free guidance (CFG) is the default mechanism for conditional generation in diffusion models, but the distribution sampled by its deterministic guided dynamics is not captured by the usual product-distribution heuristic $p_0^\omega q_0^{1-\om...

📖 Read original article


51. The World Model Remembers, the Actor Forgets: Dream Rehearsal for Continual Model-Based RL ​

Author: Gurp Nijjer
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19749v1 Announce Type: new Abstract: Model-based reinforcement-learning agents of the DreamerV3 family forget catastrophically when trained on task sequences, even when an unbounded replay buffer preserves every earlier experience. We ask a question the continual-RL literature has assumed...

📖 Read original article


52. Convergence-Latency-Aware Adaptive Modulation and Resource Allocation in RIS-Assisted Wireless Federated Learning ​

Author: Liwei Wang, Wen Chen, Jun Li, Qingqing Wu, Ming Ding, Xusheng Zhu, Qiong Wu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT

arXiv:2607.19759v1 Announce Type: new Abstract: Federated learning (FL) over wireless networks suffers from significant training latency and degraded convergence due to unreliable wireless transmission, especially under blocked propagation environments. Although reconfigurable intelligent surfaces (...

📖 Read original article


53. AlphaRoute: Large Language Models as Semantic Optimizers for Multi-Objective Routing ​

Author: Kabir Murjani, Mishri Bhavsar, Manish I. Patel, Jonti Talukdar
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AR

arXiv:2607.19768v1 Announce Type: new Abstract: Very Large Scale Integration (VLSI) global routing is an NP-hard combinatorial optimization problem requiring signal net assignment across capacity-constrained 3D grids while minimizing congestion, wirelength, and via transitions. Because traditional h...

📖 Read original article


54. An Isotropy-Preserving Spectral Cap for Muon: Theory and Three Case Studies ​

Author: Jiachun Li
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19771v1 Announce Type: new Abstract: Muon and related matrix-sign optimizers are increasingly used to pre-train large language models, but their effect on the internal geometry of individual weight matrices is not well understood. This preliminary report proposes a unified framework built...

📖 Read original article


55. OPIUM: Mitigating Steering Externalities and Over-Refusal via Dual Objective Latent Optimization ​

Author: Kavin Aravindan, Arihant Rastogi, Krishak Aneja, Aadi Prasad, Saiyam Jain, Vaishnavi Shivkumar, Ponnurangam Kumaraguru
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19806v2 Announce Type: new Abstract: Activation steering provides a lightweight mechanism for controlling large language models at inference time, but steering vectors can have unintended externalities: utility vectors may weaken safety behavior, while refusal vectors may induce over-refu...

📖 Read original article


56. Auto-Fill: Learning to Predict Missing Values Accurately with Specialist Language Models ​

Author: Yurong Liu, Yeye He, Haoyu Dong, Junjie Xing, Shi Han, Dongmei Zhang, Surajit Chaudhuri
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.DB

arXiv:2607.19847v1 Announce Type: new Abstract: Predicting missing cell values in tabular data is a fundamental problem in data cleaning. While state-of-the-art reasoning models show great promise in predicting missing values in tables, by reasoning holistically across rows and columns, they are cos...

📖 Read original article


57. Asymptotically Optimal Regret for Reinforcement Learning without Horizon Dependence ​

Author: Runlong Zhou, Zihan Zhang, Maryam Fazel, Simon S. Du
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.19854v1 Announce Type: new Abstract: We study horizon-free regret minimization for finite-horizon time-homogeneous tabular Markov decision processes with $S$ states, $A$ actions, horizon $H$, and per-trajectory total reward bounded by $1$. We propose a new algorithm and prove a regret upp...

📖 Read original article


58. Adversarial Frontiers: Minimum-Norm Attack Ensembles for Robustness Evaluation ​

Author: Luca Scionis, Luca Melis, Maura Pintor, Fabio Brau, Ambra Demontis, Giorgio Fumera, Fabio Roli, Battista Biggio
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2607.19855v1 Announce Type: new Abstract: Adversarial robustness is commonly evaluated with predefined attack ensembles, such as AutoAttack, at a single perturbation budget $\varepsilon$ and on a selective choice of perturbation norms. We argue this formulation is fundamentally limited. First,...

📖 Read original article


59. Local Causal Structure Learning in the Presence of Latent Variables and Selection Bias ​

Author: Zheng Li, Hao Zhang, Ruxin Wang, Ruichu Cai, Kun Zhang, Feng Xie
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19866v1 Announce Type: new Abstract: Discovering the direct causes and effects of a target variable from observational data is a fundamental problem in causal discovery, with broad applications in domains such as gene regulatory analysis and biomedical research. Existing causal discovery ...

📖 Read original article


60. Nonlinear Bias-Compensated Adaptive Filter and Its Application for Time-Series Prediction ​

Author: Yi Peng, Haiquan Zhao, Jinhui Hu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, eess.AS

arXiv:2607.19902v1 Announce Type: new Abstract: Most existing nonlinear adaptive filtering algorithms only account for output noise, neglecting the fact that input noise is also prevalent in practice. Although the recently proposed bias-compensated kernel least mean square (BCKLMS) algorithm address...

📖 Read original article


61. Time Series Network Utilization KPI Forecasting Using Advanced AI/ML Models ​

Author: Niraj Gadhe, Kirti Bhardwaj, Moulik Jain, Shubhi Sharma, Vinay Saini
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19974v1 Announce Type: new Abstract: The rapid proliferation of data-intensive applications, cloud infrastructure, and IoT ecosystems has made proactive resource provisioning critical for maintaining optimal network performance. However, network administrators face a constant battle again...

📖 Read original article


62. Good Practice Guide for quantifying uncertainties for machine learning models applied to photoplethysmography signals ​

Author: P. Harris, C. Bench, M. Rinkevi\v{c}ius, V. Marozas, L. Coquelin, A. Thompson, M. Nandi, U. Hackstein, P. J. Aston
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19999v1 Announce Type: new Abstract: This Good Practice Guide presents work done in the QUMPHY project (Uncertainty quantification for machine learning models applied to photoplethysmography signals) that considered both machine learning and uncertainty quantification for problems which u...

📖 Read original article


63. Post-Training in Time Series Foundation Models: A Unifying Framework ​

Author: Shifeng Xie, Ambroise Odonnat, Zehao Xiao, Lei Zan, Malik Tiomoko, Lujia Pan, Themis Palpanas, Boris N. Oreshkin, Chenghao Liu, Keli Zhang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20002v1 Announce Type: new Abstract: Time series foundation models (TSFMs) have emerged as general-purpose models for time series analysis, but pretraining alone is often insufficient for reliable downstream deployment. Bridging this gap requires further intervention to handle domain shif...

📖 Read original article


64. Generalized Kalman filter based temporal difference reinforcement learning ​

Author: Vasos Arnaoutis, Eric Lutters, Bojana Rosi'c
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CE

arXiv:2607.20010v2 Announce Type: new Abstract: In this paper, we present a generalized temporal-difference (TD) reinforcement learning framework based on the theory of conditional expectations. The value and action-value (Q-value) functions are treated as uncertain quantities, and their estimation ...

📖 Read original article


65. Zero-Shot Heart Rate Variability Forecasting from Consumer Wearables Using Time Series Foundation Models ​

Author: Luukas Per"akyl"a, Fahad Sohrab, Ville Hautam"aki, Merja Hein"aniemi, Sui Huang, Pekka Abrahamsson
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20027v1 Announce Type: new Abstract: Short-term Heart Rate Variability (HRV) forecasting could provide clinicians with actionable lead time for detecting autonomic dysfunction and adverse cardiac events. Consumer wearable devices generate fragmented, artifact-rich HRV signals that challen...

📖 Read original article


66. Test Case Prioritization for DNNs via Neural Collapse Instability ​

Author: Chunyu Liu, Mingyuan Li, Yang Li, Wenmin Li, Fei Gao, Tengfei Tu, Su-Juan Qin
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE

arXiv:2607.20046v1 Announce Type: new Abstract: With the widespread deployment of deep neural networks (DNNs) in safety-critical domains, reducing the cost of model validation under limited testing budgets has become increasingly important. Existing test case prioritization techniques often rely on ...

📖 Read original article


67. Evaluating and Mitigating Gender Bias in Pre-trained Embeddings for ML-based Recruitment ​

Author: Farnaz Faramarzi Lighvan, Lynn Houthuys
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20073v1 Announce Type: new Abstract: AI-based recruitment systems that rely on machine learning models trained on historical CV data, risk perpetuating and amplifying social biases. A key challenge arises in unstructured CV text, where pre-trained language model embeddings may infer sensi...

📖 Read original article


68. Co-Evolving LLM Evaluators and Policies via DynamicRubric ​

Author: Beining Wang, Weihang Su, Hongtao Tian, Hao Kong, Tao Yang, Ting Yao, Qingyi Pan, Yueyue Wu, Qingyao Ai, Min Zhang, Yiqun Liu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20083v2 Announce Type: new Abstract: Post-training with evaluator feedback on policy-induced samples serves as a major mechanism for improving large language models. As policies improve, these sampled responses become close in quality. These close candidates create a bottleneck for policy...

📖 Read original article


69. Autonomous Collaborative Learning Among an Ensemble of Tsetlin Machines with Consensus-Based Inference ​

Author: Yehuda Rudin, Osnat Keren, Michal Yemini, Alexander Fish
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.MA

arXiv:2607.20124v1 Announce Type: new Abstract: Tsetlin Machine (TM) is a rule-based machine-learning algorithm comprising collectives of two-action Tsetlin Automata (TAs) that cooperatively form conjunctive logical clauses from Boolean inputs through stochastic feedback. Although few recent studies...

📖 Read original article


70. CURED: Creating, Understanding, and Repairing Errors Demonstrator ​

Author: Nicholas Chandler, Sebastian J"ager, Philipp Jung, Felix Bie{\ss}mann
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20140v1 Announce Type: new Abstract: Detecting and cleaning errors in tabular data is a prerequisite for data intense software applications. Recent research at the intersection of Machine Learning (ML) and Database Management Systems (DBMS) highlights the potential of statistical learning...

📖 Read original article


71. Active Inference as a Convex Markov Decision Process ​

Author: Nikola Milosevic, Nicol'as Hinrichs, Nico Scherf
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.20152v1 Announce Type: new Abstract: Active Inference (AIF) frames adaptive behavior as the minimization of expected free energy (EFE), combining epistemic and pragmatic objectives within a single variational principle. We frame AIF as policy optimization and show that, for closed-loop co...

📖 Read original article


72. Local Stability and Gaussian Smoothing of Quantized Neural Networks ​

Author: Sergey Salishev, Anton Makarov, Oleg Granichin
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY, math.OC

arXiv:2607.20153v1 Announce Type: new Abstract: We study Gaussian averaging as a smooth surrogate for quantized neural models. Under bounded local oscillation, we derive a local dimension-dependent bound on |f-g|, linking Gaussian smoothing to the stability analysis of discontinuous networks. We com...

📖 Read original article


73. Self-organizing Architecture of Receptron Units: a Hardware-Aware Framework for Edge Intelligence ​

Author: Stefano Radice, Ludovico Casaccia, Riccaro Emanuele Beccalli, Bruno Paroli, Paolo Milani
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.ET

arXiv:2607.20162v1 Announce Type: new Abstract: The growing demand for intelligent processing at the edge of IoT networks is constrained by the severe computational and memory limitations of microcontroller units, which render impractical conventional deep learning approaches. We propose a neuromorp...

📖 Read original article


74. Instance Hardness-Based Relevance for Imbalanced Regression ​

Author: Vitor M. Leitao, Juscimara G. Avelino, George D. C. Cavalcanti, Rafael M. O. Cruz
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20173v2 Announce Type: new Abstract: Imbalanced regression problems arise when the target variable has an asymmetric distribution, resulting in underrepresented value ranges in the dataset. Traditional approaches for identifying rare instances rely on a relevance function that assigns hig...

📖 Read original article


75. On Optimization Complexity of Second-Order Certified Unlearning ​

Author: Nikita Doikov, Anastasia Koloskova
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2607.20192v1 Announce Type: new Abstract: We study machine unlearning: the removal of memorized training data from a trained model. Specifically, we investigate the algorithmic complexity of certified unlearning from an optimization perspective. We formalize the goal of an unlearning algorithm...

📖 Read original article


76. OLEDLM: A Unified Language Model for OLED Molecular Design ​

Author: Fukang Wen, Yuchong Tang, Jingyuan Li, Beichen Wang, Yixuan Jiang, Xiaoyi Jiang, Yaxuan Liu, Shunyu Wang, Zuoqiang Shi, Yi Zhu, Yanan Zhu, Pipi Hu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20194v1 Announce Type: new Abstract: The development of organic light-emitting diode (OLED) materials faces the compounded challenges of an astronomically large chemical space, stringent quantum-chemical constraints, and a scarcity of labeled data. Although the question of OLED generation...

📖 Read original article


77. The Quadrilateral Loss: Additivity as a Measurable Behavior of Dense Neural Networks ​

Author: Antonio Di Cecco
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20201v1 Announce Type: new Abstract: Additive models buy interpretability by forbidding feature interactions, a constraint that neural instantiations enforce architecturally. We introduce the quadrilateral loss, a differentiable penalty that treats additivity as a measurable behavior inst...

📖 Read original article


78. ELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training Transformers ​

Author: Mahdi Heidari, Mohammad Mahdi Rahimi, Jaekyun Moon
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20214v1 Announce Type: new Abstract: The quadratic $N\times N$ attention score matrix remains a central obstacle to extending Transformers to longer input lengths. Existing efficient attention methods usually reduce this bottleneck by either imposing sparsity, so that each query attends t...

📖 Read original article


79. User-Centric Modeling of Transactional Sequences with Explainable State Space Models ​

Author: Ivan Palagin
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20228v1 Announce Type: new Abstract: We propose a hybrid approach for user-centric modeling of transactional event sequences that combines contrastive representation learning (CoLES) with State Space Models (SSMs). While contrastive methods yield high-quality compressed user representatio...

📖 Read original article


80. PIER: Physics-Informed Environmental Retrieval for Time-Series Modeling ​

Author: Shiyuan Luo, Runlong Yu, Chonghao Qiu, Yue Qin, Rahul Ghosh, Robert Ladwig, Paul C. Hanson, Yiqun Xie, Xiaowei Jia
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20230v1 Announce Type: new Abstract: Accurate modeling of environmental systems is fundamental to scientific understanding and decision-making, yet remains challenging because observations are limited and physical dynamics vary across systems. Retrieval-augmented approaches offer a natura...

📖 Read original article


81. PhaseAware: Interpretable Human-in-the-Loop Rehabilitation Scoring with Boundary Monitoring ​

Author: Yankai Zheng, Yuhe Liu, Yuxin Ma, Tianci Xue, Jiayuan Tian, Yu Fu, Yuxuan Hu, Jianing Wang, Zichun Xiao, Junya Mu, Shaohui Ma
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20237v1 Announce Type: new Abstract: Rehabilitation scoring systems are most useful when their outputs can be reviewed and interpreted within clinical workflows. This study presents PhaseAware, a compact framework for continuous rehabilitation quality assessment that combines a temporal b...

📖 Read original article


82. Breaking the $T^{3/4}$ Barrier for Regret Minimization With Bi-Dimensional CDFs ​

Author: Matteo Castiglioni, Anna Lunghi, Alberto Marchesi
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20258v1 Announce Type: new Abstract: We study regret minimization for learning CDF-related objectives of the form [ g(x)\cdot\mathbb{P}_{X\sim\mathcal{D}}(X\le x), ] over $[0,1]^2$, where $g$ is a known Lipschitz function and $\mathcal{D}$ is an unknown distribution. At each round $t$, ...

📖 Read original article


83. Interpretable Fuzzy Rule-Based Regression Extension for Ex-Fuzzy Library ​

Author: Cayan Deniz Kucuktopana, Javier Fumanal-Idocin, Richard Pitts, Javier Andreu-Perez
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20277v1 Announce Type: new Abstract: Machine learning models achieve high predictive accuracy in regression tasks, but their deployment in safety-critical and regulated domains requires interpretability. While fuzzy rule-based systems offer transparent, linguistically explicit interpretab...

📖 Read original article


84. The Blessing of Dimensionality: How Near-Orthogonality in High-Dimensional Spaces Explains Temporal Portability ​

Author: Abigail Woodring, Adrian Chan, Rana Muhammad Shahroz Khan, Sukwon Yun, Chau-Wai Wong, Tianlong Chen
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.20301v1 Announce Type: new Abstract: Fine-tuning has been widely used to adapt large language models (LLMs) for domain-specific tasks. Parameter efficient fine-tuning (PEFT) methods such as low-rank adaptation (LoRA) are frequently used to reduce computational costs. PortLLM is a training...

📖 Read original article


85. Classical Hardware Acceleration of Quantum Autoencoders for Real-Time Anomaly Detection in Collider Experiments ​

Author: Ivan Ge, Sagar Addepalli, Abhilasha Dave, Julia Gonski
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, hep-ph, physics.ins-det

arXiv:2607.20302v1 Announce Type: new Abstract: Quantum machine learning (QML) algorithms in high energy physics (HEP) can efficiently represent and leverage long-range, high-order correlations in high-dimensional collider data, potentially with fewer parameters and favorable scaling relative to cla...

📖 Read original article


86. Multi-modal transformer for signal classification in nanopore blockade experiments ​

Author: Sandro Kuppel, Julian Ho{\ss}bach, Samuel Tovey, Christian Holm
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph, q-bio.BM

arXiv:2607.20323v1 Announce Type: new Abstract: Nanopore devices have emerged as powerful tools for single-molecule sensing, with potential for rapid, portable diagnostics. They detect changes in ionic current as analytes enter nanometer-scale pores, providing a means of identifying diverse biomarke...

📖 Read original article


87. Interval and fuzzy physics-augmented neural networks (iPANN and fPANN) for uncertainty quantification and propagation in constitutive modeling ​

Author: Somesh Pratap Singh, Govinda Anantha Padmanabha, Jingye Tan, Steven Yang, Reese E. Jones, D. Thomas Seidl, Nikolaos Bouklas
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph

arXiv:2607.20339v1 Announce Type: new Abstract: Constitutive modeling under uncertainty remains a central challenge for reliable mechanics simulations, particularly when the available stress-deformation data are sparse, noisy, or heterogeneous. We propose interval and fuzzy physics-augmented neural ...

📖 Read original article


88. Variance-reduced Domain Adaptation using Paired Sampling ​

Author: Andrea Napoli
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20367v1 Announce Type: new Abstract: Correlation alignment and the maximum mean discrepancy are two widely used distribution-matching frameworks for unsupervised domain adaptation (UDA). However, high variance in these losses has been shown to undermine their effectiveness in minibatch op...

📖 Read original article


89. Online Variance Reduction for Domain Adaptation on Streaming Data ​

Author: Andrea Napoli
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20374v1 Announce Type: new Abstract: This paper studies the problem of stochastic variance reduction (SVR) for the maximum mean discrepancy (MMD) and correlation alignment (CORAL) loss functions. Although various offline SVR algorithms for these losses have been proposed, these are incomp...

📖 Read original article


90. PG-KINN: A Physics-Informed Petrov-Galerkin Kolmogorov-Arnold Network for Solving Forward and Inverse PDEs ​

Author: Amirhossein Sadr, Nima Soltani, Vahideh Moghtadaiee, Aida Pakniyat, Dara Rahmati, Saeid Gorgin
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2607.20378v1 Announce Type: new Abstract: Physics-informed learning of partial differential equations (PDEs) has been dominated by multilayer perceptrons (MLPs), whose spectral bias and dense parameterization limit both accuracy and interpretability. Kolmogorov Arnold Networks (KANs) mitigate ...

📖 Read original article


91. Isaac Sim-to-Real: Reinforcement Learning based Locomotion for Quadrupeds ​

Author: Jordan Dowdy, Jean Chagas Vaz
Published: 7/23/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.18135v1 Announce Type: cross Abstract: Learning-based approaches to locomotion have risen in popularity in recent years, showing the capability for complex legged locomotion and whole-body control. Reinforcement learning (RL), the primary learning-based approach for locomotion, often util...

📖 Read original article


92. Towards Torque-Driven Reinforcement Learning for Quadruped Locomotion ​

Author: Jordan Dowdy, Jean Chagas Vaz
Published: 7/23/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SY, eess.SY

arXiv:2607.18365v1 Announce Type: cross Abstract: Reinforcement learning (RL) for legged robots is advancing locomotion, demonstrating its ability to adapt to new and challenging terrain. Traditionally, these RL locomotion frameworks are position-based, making the policy less adaptable to terrain ty...

📖 Read original article


93. Hybrid LSTM-Graph Neural Framework for Robust Financial Fraud Detection and Adversarial Resilience ​

Author: Mariam Zakaria Moussa Ali
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.19350v1 Announce Type: cross Abstract: Financial institutions face significant challenges in detecting sophisticated money laundering patterns, such as smurfing and layering, due to extreme data imbalance (0.13% fraud rate) and evolving adversarial evasion tactics. This paper proposes Fra...

📖 Read original article


94. Benchmarking Confidential GPU Inference on NVIDIA H100 under Intel TDX ​

Author: Wei Wang, Abdul Hyee Waqas, Burns Smith
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.19353v1 Announce Type: cross Abstract: Confidential computing is becoming a practical deployment requirement for AI inference workloads that process sensitive inputs or protect proprietary model assets. However, the performance cost of enabling confidential execution for GPU-accelerated l...

📖 Read original article


95. Stochastic Primal-Dual Decoding for Multiobjective Generative Recommender Systems ​

Author: Dmitrii Moor, Ben Carterette, Senthilkumar Krishnamoorthy, Kyle Kretschman, Denis Beslic, Melissa Yalla, Alice Y Wang, Mounia Lalmas
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.19357v1 Announce Type: cross Abstract: Recent advances in recommender systems (RS) have shown substantial performance gains through generative modelling. In practice, recommendation often involves constructing slates -- ordered lists of items -- that must satisfy multiple objectives beyon...

📖 Read original article


96. Mitigating Scaffolding Collapse in Socratic Tutors via Representation Alignment ​

Author: Jing Shao, Qifeng Wu, Hanyu Zhang, Sixia Sun, Jun Zhuang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.19371v1 Announce Type: cross Abstract: Large language model (LLM)-based Socratic tutors increasingly guide students through multi-turn questioning, but they can suffer from scaffolding collapse: under sustained student pressure, a tutor gradually abandons guided inquiry and reveals soluti...

📖 Read original article


97. Refnd: Preventing Data Leakage in Relational Datasets ​

Author: Anthony Lavertu, Jacob Cote, Jacques Corbeil, Sophie Gobeil, Pascal Germain
Published: 7/23/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG

arXiv:2607.19376v1 Announce Type: cross Abstract: Machine learning models trained on biochemical data are routinely evaluated using splits that fail to account for relational structure, causing information leakage and over-optimistic performance estimates. Existing splitting methods lack theoretical...

📖 Read original article


98. Reliability-Aware Hard--Soft Physics-Informed Neural Networks for Robust Learning of Challenging Partial Differential Equations ​

Author: Duc Tien Nguyen, Hang Tran, Trinh Minh Tuan, Nguyen Duc Manh, Dinh Gia Ninh
Published: 7/23/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, physics.flu-dyn

arXiv:2607.19377v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) provide a mesh-free framework for solving partial differential equations, but their training is often affected by loss imbalance, optimization stiffness, and difficulty in capturing localized or multi-mode sol...

📖 Read original article


99. Trend strength predicts when generative foundation models win: a power-controlled benchmark, a mechanism, and an actionable selection rule ​

Author: Ahmed Cherif
Published: 7/23/2026, 4:00:00 AM
Categories: stat.AP, cs.LG

arXiv:2607.19383v1 Announce Type: cross Abstract: Pretrained generative foundation models cast forecasting as conditional generation from a learned predictive distribution and forecast unseen series zero-shot. We establish three results that turn their reported success into an actionable, mechanisti...

📖 Read original article


100. Simulating Eutopia: Revisiting Long-term Fairness with Outcomes, Performativity, and Dynamics ​

Author: Vedant Palit, Udvas Das, Brahim Driss, Debabrota Basu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG, stat.ML

arXiv:2607.19389v1 Announce Type: cross Abstract: As AI-driven Decision Makers (ADMs) influence our socioeconomic reality, their roles in both enhancing efficiency and amplifying the social biases have drawn attention. In this paper, we revisit the nuances of long-term `fairness' achievable by an AD...

📖 Read original article


101. Making Single-Cell Data Distillation Auditable: Traceable Real-Cell Coresets via Discrete Min-Max Selection ​

Author: Yaodi Luo, Peize He, Bowen Han, Lingbei Mengg
Published: 7/23/2026, 4:00:00 AM
Categories: q-bio.GN, cs.AI, cs.LG

arXiv:2607.19426v1 Announce Type: cross Abstract: Single-cell datasets are increasingly costly to store, audit, and reuse for model training. Dimensionality reduction and dataset distillation can reduce this burden, but conventional distillation methods often produce synthetic expression profiles th...

📖 Read original article


102. BaseRT: Advancing Best-in-Class LLM Inference with Apple M5 Neural Accelerators ​

Author: Fabian Waschkowski, Prabod Rathnayaka, Lukas Wesemann
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AR, cs.AI, cs.CL, cs.DC, cs.LG, cs.PF

arXiv:2607.19438v1 Announce Type: cross Abstract: Apple's M5 generation introduces a redesigned GPU architecture in which every core carries a dedicated Neural Accelerator: on-die matrix units exposed through the Metal~4 tensor API. We show that BaseRT, our native Metal inference runtime for large l...

📖 Read original article


103. A Bayesian Framework for Built-in Input Dimension Reduction for Gaussian Process Modeling ​

Author: Eric Herrison Gyamfi, Emily L. Kang, Bledar A. Konomi, Guang Lin
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR, stat.AP, stat.CO, stat.ME

arXiv:2607.19498v1 Announce Type: cross Abstract: Gaussian process (GP) modeling is widely used in computational science and engineering. However, fitting a GP to high-dimensional inputs remains challenging due to the curse of dimensionality. While various methods have been proposed to reduce input ...

📖 Read original article


104. Tensor Network Machine Learning for Wildfire Susceptibility Mapping: from Grokking Dynamics to Quantum Mixedness of Class Representations ​

Author: Domenico Pomarico, Alessandra Costantino, Gabriel Ramirez Sanchez, Loredana Bellantuono, Davide D' Al`o, Mario Elia, Alessandro Fania, Francesco Giordano, Niloofar Kheirkhahan, Raffaele Lafortezza, Ester Pantaleo, Sabina Tangaro, Roberto Bellotti, Alfonso Monaco, Nicola Amoroso
Published: 7/23/2026, 4:00:00 AM
Categories: physics.soc-ph, cs.LG, physics.data-an

arXiv:2607.19503v1 Announce Type: cross Abstract: A quantum-inspired tensor network framework for wildfire susceptibility classification in the Gargano region is introduced, leveraging AlphaEarth embeddings and Matrix Product State models. The approach combines scalable geospatial representations wi...

📖 Read original article


105. Boltzmann-Expected Molecular Design with Decoupled Annealing Flows ​

Author: Selma Moqvist, Richard Beckmann, Ross Irwin, Roc'io Mercado, Simon Olsson
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.19519v1 Announce Type: cross Abstract: Most 3D properties relevant to molecular design, including free energies and shape descriptors, are $\textit{expectations}$ over the Boltzmann distribution over 3D configurations of a molecular graph. However, existing property-guided generative mode...

📖 Read original article


106. Equilibrium Causal Games: Separation, Identification, and the Identifiability of Cyclic Latent States ​

Author: Faraz Dadgostari, Neda Nazemi
Published: 7/23/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.SY, eess.SY

arXiv:2607.19531v1 Announce Type: cross Abstract: Power grids, markets, and interacting populations, settle into feedback driven equilibria observed through unknown sensors. Our Equilibrium Causal Game (ECG) joins a game to its cyclic causal model, hidden inputs, sensor map, and rules for interventi...

📖 Read original article


107. RELTA-SGLD: Relative-Growth Localized Taming for Nonconvex Stochastic-Gradient Langevin Learning ​

Author: Yiwei Zhou, Ziheng Chen
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC

arXiv:2607.19544v1 Announce Type: cross Abstract: We introduce RELTA-SGLD, a taming scheme that stabilizes superlinear stochastic-gradient updates while reducing unnecessary suppression of the original learning drift. A threshold determines where the taming turns on, while a relative-growth principl...

📖 Read original article


108. Online Optimization of Difference-of-Convex Compositions with Smooth Mappings ​

Author: Jingwei Ji, Jong-Shi Pang, Renyuan Xu
Published: 7/23/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2607.19553v1 Announce Type: cross Abstract: We study online optimization for a broad class of structured non-convex non-smooth problems where each loss is a composition of a difference-of-convex function with a smooth mapping, and the feasible region is defined by constraint functions of the s...

📖 Read original article


109. Machine-learned syndrome post-selection for reliable quantum error correction ​

Author: Tobias Haug, Askery Canabarro, Leandro Aolita
Published: 7/23/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.19563v1 Announce Type: cross Abstract: Quantum error correction can be enhanced by post-selecting out runs that are likely to produce a logical failure, but the most accurate measures for that require costly decoder-level information. We introduce a practical, decoder-agnostic post-select...

📖 Read original article


110. On the Computational Complexity of Structural Generalization ​

Author: Zichao Wei
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.19573v1 Announce Type: cross Abstract: Structural generalization has been measured repeatedly by several benchmarks, yet it has never been formally defined. We give a definition that translates the two premises (compositional structure and unbounded generalization) into mathematical langu...

📖 Read original article


111. Knowledge-Centric Self-Improvement ​

Author: Xuefei Julie Wang, Lauren Hyoseo Yoon, Chengrui Qu, Amanda Zichang Wang, Atharva Sehgal, Eric Mazumdar, Yisong Yue
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2607.19592v1 Announce Type: cross Abstract: Self-improving AI systems typically treat the agent as the object that improves, by optimizing prompts, workflows, harnesses, or even the agent's own code. This agent-centric view can make improvements expensive to maintain and difficult to transfer,...

📖 Read original article


112. A Deep Learning Framework for Predicting Solar EUV Irradiance During Significant Flares ​

Author: Sathvik Soman, Jason T. L. Wang, Haimin Wang, Haodi Jiang
Published: 7/23/2026, 4:00:00 AM
Categories: astro-ph.SR, astro-ph.IM, cs.LG

arXiv:2607.19597v1 Announce Type: cross Abstract: We present FlareEUV, a multimodal deep learning framework for predicting daily extreme ultraviolet (EUV) irradiance at 6.5 nm over three consecutive days during significant solar flares, using multi-instrument observations from NASA's Solar Dynamics ...

📖 Read original article


113. Deep Shape Regression for Planar Curves with Multimodal Covariates ​

Author: Manuel Pfeuffer, Roshan Prakash Rane, Hadya Yassin, Kerstin Ritter, Sonja Greven
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ME, cs.CV, cs.LG, q-bio.QM, stat.ML

arXiv:2607.19600v1 Announce Type: cross Abstract: The shape of a planar curve is the geometric information that remains once translation, rotation, scale and reparametrisation are removed and is of interest in many health applications, e.g. in neuroimaging. We propose a deep shape regression model f...

📖 Read original article


114. Scaling Laws for Hypernetwork-Based Knowledge Injection in Large Language Models ​

Author: Nischay Dhankhar, Dos Baha, Abulhair Saparov
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.19604v1 Announce Type: cross Abstract: Injecting factual knowledge into large language models (LLMs) reliably and at scale remains an open challenge. Hypernetworks provide a promising solution to large-scale knowledge injection. Although hypernetworks are typically applied for test-time a...

📖 Read original article


115. CRB-Driven Beamforming and Trajectory Optimization for UAV-assisted ISAC System ​

Author: Yi Yang, Qianqian Zhang, Huaxia Wang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, cs.RO, math.IT

arXiv:2607.19609v1 Announce Type: cross Abstract: In this paper, we study an unmanned aerial vehicle (UAV)-assisted integrated sensing and communication (ISAC) system, where a UAV enhances the sensing capability of a base station (BS) towards a target while ensuring reliable communication towards a ...

📖 Read original article


116. Causal dictionary learning reveals and validates transcription-factor binding features in genomic language models ​

Author: Sarwan Ali
Published: 7/23/2026, 4:00:00 AM
Categories: q-bio.GN, cs.AI, cs.LG

arXiv:2607.19618v1 Announce Type: cross Abstract: Genomic language models achieve strong performance across regulatory-genomics tasks, yet what these models internally represent remains opaque, and the field lacks a principled procedure for verifying that an apparent ``concept'' inside a model is re...

📖 Read original article


117. From Bit-Position Sensitivity to Unequal Error Protection for DNN Inference Memory ​

Author: Muhammad Husnain Mubarik, Karthik Mohan Kumar, Pedro Antonio Pena, Keshavan Varadarajan, Kunal Tyagi
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AR, cs.LG

arXiv:2607.19623v1 Announce Type: cross Abstract: We characterize per-bit-position fault sensitivity in ML inference across 16 workloads -- spanning transformer-based models and attention-free CNNs -- and across three floating-point formats. Our central empirical finding is a sharp bit-sensitivity t...

📖 Read original article


118. Leveraging ECRAM for Edge Continual Learning ​

Author: Nabila Tasnim, Haoran Liu, Qing Cao, Saugata Ghose
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AR, cs.ET, cs.LG

arXiv:2607.19661v1 Announce Type: cross Abstract: Several edge computing platforms, such as autonomous vehicles and smart sensing devices, need to adapt to dynamic environments in real time by learning from new data in the field. Continual learning has emerged as a promising solution for edge traini...

📖 Read original article


119. FedLSG: LLM-Enhanced Semantic Calibration for Federated Graph Backdoor Defense ​

Author: Chenyu Zhou, Yabin Peng, Wei Huang, Kunlin Li, Shuaishuai Zhang, Xinyuan Miao
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2607.19674v1 Announce Type: cross Abstract: Federated Graph Neural Networks (FedGNNs) are highly vulnerable to backdoor poisoning, yet existing defenses typically rely on rule-based approaches that lack semantic understanding, making them vulnerable to stealthy triggers and harmful to benign s...

📖 Read original article


120. Reference-Free Evaluation of Reasoning in Open-Ended Question Answering ​

Author: Guneet Singh Kohli, Yuxiang Zhou, Michael Sejr Schlichtkrull, Gregory E Dean, Maria Liakata
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.19678v1 Announce Type: cross Abstract: AI-generated answers in high-stakes domains are often fluent but difficult to verify, especially when they contain multi-step reasoning rather than a single final answer. We propose a reasoning-based, reference-free framework for auditing LLM-generat...

📖 Read original article


121. Nuclear Quantum Effects as a Denoising Problem ​

Author: Weizhou Wang, Jonathan Weare, Aaron R. Dinner
Published: 7/23/2026, 4:00:00 AM
Categories: physics.chem-ph, cond-mat.stat-mech, cs.LG, physics.comp-ph, quant-ph

arXiv:2607.19680v1 Announce Type: cross Abstract: Nuclear quantum effects are rigorously captured by imaginary-time path integrals, which map the quantum Boltzmann distribution onto a ring polymer of classical replicas. Yet the nuclear masses, the coupling to the environment, and the boundary condit...

📖 Read original article


122. Multi-Mask Diffusion Language Models for Few-Step Generation ​

Author: Sijin Chen, Yinuo Ren, Heyang Zhao, Ziheng Cheng, Quanquan Gu, Lexing Ying
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.19686v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) are a promising family of language generators, but achieving high-quality few-step generation remains challenging. In MDMs, all forward trajectories collapse to a single fully masked state, leaving no terminal entropy f...

📖 Read original article


123. Optimal Recalibration of an Online Predictor ​

Author: Lunjia Hu, Kevin Tian, Chutong Yang
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.DS, cs.LG

arXiv:2607.19689v1 Announce Type: cross Abstract: We study the problem of recalibrating an online predictor [KE17, OKS24]: given an arbitrary "hint" sequence of forecasts, the learner must output new predictions that are calibrated while incurring small excess error relative to the original forecast...

📖 Read original article


124. SLPO: Scaling Latent Reasoning via a Surrogate Policy ​

Author: Runyang You, Zhiyuan Liu, Yongqi Li, Wenjie Li
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.19691v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards has become the predominant recipe for eliciting test-time scaling in explicit Chain-of-Thought reasoners. Yet this scaling path remains computationally costly, since every intermediate step must be decod...

📖 Read original article


125. Data-Poisoning Audits for Causal Effect Estimation ​

Author: Kwangho Kim
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2607.19692v1 Announce Type: cross Abstract: Observational causal analyses increasingly pool records across sites, vendors, and collection systems, creating vulnerability to append-only attacks in which plausible records are strategically selected to alter a reported treatment effect. We develo...

📖 Read original article


126. Domain-Adapted Power Curve for Cross-Farm Applications ​

Author: Ahmadreza Chokhachian, V. Roshan Joseph, Yu Ding
Published: 7/23/2026, 4:00:00 AM
Categories: stat.AP, cs.LG

arXiv:2607.19744v1 Announce Type: cross Abstract: The wind energy industry relies on accurate power curve models to make power forecast, evaluate turbine performance, quantify upgrade, or support site-planning decisions. In this paper, we focus on site-planning power curves, i.e., we investigate how...

📖 Read original article


127. Machine Can Automatically Discover Parametric Functions to Model HEP Data ​

Author: Ho Fung Tsoi, Dylan Rankin, Cecile Caillol, Miles Cranmer, Sridhara Dasu, Javier Duarte, Philip Harris, Elliot Lipeles
Published: 7/23/2026, 4:00:00 AM
Categories: hep-ex, cs.LG

arXiv:2607.19750v1 Announce Type: cross Abstract: In HEP data analyses, finding an adequate function to model binned data has largely relied on a manual process: guess a functional form by intuition, fit, examine, then repeat until successful. We show that this iterative process can be automated by ...

📖 Read original article


128. Learning the Arabic Dialect Continuum as a Continuous Space: A Regression Approach to Speaker Origin Prediction ​

Author: Mohamed Aziz Khadraoui, Adel Ammar, Bilel Benjdira, Zahid Khan, Skander Turki, Wadii Boulila
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG, cs.NE

arXiv:2607.19751v1 Announce Type: cross Abstract: We present a regression-based approach to Arabic dialect geolocation that models dialectal variation as a continuous geographic space rather than discrete categories. Speaker origin is predicted as continuous latitude-longitude coordinates using a hi...

📖 Read original article


129. A Multiclass Quantum Aligned Centroid Kernel ​

Author: Kilian Tscharke, Pascal Debus
Published: 7/23/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.19782v1 Announce Type: cross Abstract: Kernel methods are powerful tools in machine learning but commonly used full-Gram kernels face three key limitations: (1) quadratic scaling with training set size; (2) the use of fixed, non-trainable kernels; and (3) the absence of an intrinsic formu...

📖 Read original article


130. A Structure-Adaptive Random Feature Method for High-Dimensional Elliptic PDEs ​

Author: Jiale Linghu, Hao Dong, Yangshuai Wang
Published: 7/23/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA

arXiv:2607.19786v1 Announce Type: cross Abstract: Random-feature methods reduce high-dimensional elliptic PDE collocation to linear coefficient problems, but full-dimensional trial spaces overlook lower-dimensional structure. We introduce the Hierarchical Analysis-of-Variance Random Feature Method (...

📖 Read original article


131. TriAgent: Divergence-Aware Multi-Agent Committees for Cost-Efficient Financial Sentiment Analysis ​

Author: Isabel Xu (The Overlake School), Cynthia Xu (The Overlake School), Rachel Ren (Edwards Vacuum Inc.), Cong Guo (The University of Memphis), Jiacheng Ding (The University of Memphis)
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CL, cs.CE, cs.DB, cs.LG

arXiv:2607.19794v1 Announce Type: cross Abstract: Production LLM-based financial sentiment analysis faces a structural cost trap: most queries are trivially classifiable, yet expensive cloud reasoners process them all, and the bill scales linearly with user count. We present TriAgent, a multi-agent ...

📖 Read original article


132. Zero-Observation User Reactivation with Gap-Driven Dimensional Gating ​

Author: Jiandong Ding, Tianying Liu, Fuyuan Liu, Huijie Qin, Tiandeng Wu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.19802v1 Announce Type: cross Abstract: Sequential recommendation (SR) models capture continuously observed behavior, but a returning user may have no interactions for months or years. We define this setting as Zero-Observation Reactivation: the user has a pre-gap history, while the platfo...

📖 Read original article


133. Dreamer-CPC: Message Learning with World Models for Decentralized Multi-agent Reinforcement Learning ​

Author: Taisuke Takayama, Naoto Yoshida, Tadahiro Taniguchi
Published: 7/23/2026, 4:00:00 AM
Categories: cs.MA, cs.LG

arXiv:2607.19809v1 Announce Type: cross Abstract: In multi-agent reinforcement learning (MARL), inter-agent communication is effective for improving performance under partial observability. Representation learning-based approaches enable decentralized agents to learn messages grounded in their own o...

📖 Read original article


134. Hypothesis-and-Refinement Learning of Organic Structures from Multimodal Spectroscopic Data ​

Author: Chengchun Liu, Zhiyuan Yan, Li Yuan, Hao Li, Boxuan Zhao, Yonghong Tian, Bartosz A. Grzybowski, Fanyang Mo
Published: 7/23/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, physics.comp-ph, physics.data-an

arXiv:2607.19816v1 Announce Type: cross Abstract: Determining molecular structures from spectroscopic data remains fundamentally challenging because the inverse problem is intrinsically underdetermined: individual spectra are sparse, low-dimensional, and encode only partial structural evidence relat...

📖 Read original article


135. Know Your Agent: Reconnaissance-Driven Pentesting of AI Agents ​

Author: Or Zion Eliav, Eyal Lenga, Shir Bernstien, Yisroel Mirsky
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG

arXiv:2607.19837v1 Announce Type: cross Abstract: Traditional pentesting uses reconnaissance at each step to uncover unseen weaknesses, build stronger attacks, and advance the objective; we argue that AI agents require the same treatment. We formalize agent reconnaissance by modeling the process and...

📖 Read original article


136. DocOps: A Verifiable Benchmark for Autonomous Agents in Complex Document Operations ​

Author: Jiazhen Jiang, Boxi Cao, Lingyong Yan, Yaojie Lu, Hongyu Lin, Shuaiqiang Wang, Dawei Yin, Xianpei Han, Le Sun
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.19865v1 Announce Type: cross Abstract: As autonomous agents rapidly evolve, their ability to reliably manipulate ubiquitous digital documents has become critical for enabling general-purpose AI assistants and automating complex workspace workflows. In this paper, we introduce DocOps, a de...

📖 Read original article


137. Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage ​

Author: Shay Seiya McDonnell, Avantika Singh, Quoc-Viet Pham, Vratislav Havlik, Gregory M. P. O'Hare
Published: 7/23/2026, 4:00:00 AM
Categories: cs.MA, cs.LG

arXiv:2607.19899v1 Announce Type: cross Abstract: Disagreement-triggered escalation can create a structural blind spot in multi-agent arbitration: as base learners improve, they tend to converge, weakening safety monitoring where correlated failures concentrate. We term this correlated agreement bli...

📖 Read original article


138. Diffusion ReRoll: Revisable Denoising for Robotic Sequential Prediction ​

Author: Seonsoo Kim, Seongil Hong, Jun-Gill Kang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.19919v1 Announce Type: cross Abstract: We propose Diffusion ReRoll, a diffusion-based framework for robotic sequential prediction that enables revisable denoising over horizons. Existing diffusion-based sequence predictors typically perform a single monotonic denoising process. In contras...

📖 Read original article


139. HijackKV: New Threat in Position-Independent KV Cache Reuse ​

Author: Yichi Zhang, Zhiqi Wang, Huan Zhang, Yuchen Yang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2607.19957v1 Announce Type: cross Abstract: Key-Value (KV) cache reduces inference latency in large language models (LLMs). Traditional prefix-based reuse has low cache hit rates across inference requests because it requires exact token and position matches. To improve efficiency, recent syste...

📖 Read original article


140. The Giant Hippocampus: From Structural Monoculture to a System of Systems ​

Author: Jaeho Seol
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NE, q-bio.NC

arXiv:2607.19973v1 Announce Type: cross Abstract: AI researchers describe state-of-the-art models as one thing repeated at scale: the Transformer, wired identically for text, pixels, or speech. Neuroscientists describe the cortex as a mosaic - dense Layer 4 in visual cortex for spatial encoding, thi...

📖 Read original article


141. Taming the Security-Energy Paradox: A Green AI Approach to Optimized Android Malware Detection ​

Author: Shrinidhi Sridhar, Vikas K. Malviya
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2607.20003v1 Announce Type: cross Abstract: An increase in advanced Android malware requires the use of deep learning models, which can run on Android devices. But there is a trade-off between security and energy use, as strong detection models can drain the battery of devices fast. This work ...

📖 Read original article


142. PN-QNN: Harnessing Physical Noise as a Native Regularizer in Photonic Hybrid Quantum Neural Networks ​

Author: Farah Elnakhal, Alberto Marchisio, Nouhaila Innan, Gabriel Falcao, Muhammad Shafique
Published: 7/23/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.20045v1 Announce Type: cross Abstract: Physical noise in near-term quantum hardware is usually treated as a nuisance to suppress. We ask whether it can instead act as a hardware-native regularizer for photonic hybrid quantum-classical neural networks (PHQCNNs), analogous to noise-injectio...

📖 Read original article


143. Antigen-specific Antibody Multi-modal Foundation Model for Functional Antibody Design ​

Author: Xiaoliang Shi, Zichen Wang, Runze Ma, Zhongyue Zhang, Shuangjia Zheng
Published: 7/23/2026, 4:00:00 AM
Categories: q-bio.BM, cs.LG

arXiv:2607.20057v1 Announce Type: cross Abstract: Antibodies are essential proteins that play a central role in immune recognition by binding specific antigen molecules. Although recent protein language models have enabled progress in single-chain protein modeling and generation, they often fall sho...

📖 Read original article


144. Non--negative matrix factorization using the \textit{R} package \textsf ​

Author: Volkan Sevin\c{c}, Nikolas Kontemeniotis, Theodoros Perdikis, Michail Tsagris
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.20084v1 Announce Type: cross Abstract: Non--negative matrix factorization (NMF) has become an established dimensionality reduction technique for extracting latent structures from non--negative data and has found widespread applications in fields such as bioinformatics, text mining, image ...

📖 Read original article


145. Cumsum-Composable Phase Transport for Low-Cost Streaming Keyword Spotting ​

Author: Mahesh Godavarti
Published: 7/23/2026, 4:00:00 AM
Categories: cs.SD, cs.LG

arXiv:2607.20086v1 Announce Type: cross Abstract: State-space sequence models are attractive for streaming speech because they maintain compact recurrent state, but scan-style training kernels can have unfavorable constants for short audio tasks. We study cumsum-composable phase transport, a streami...

📖 Read original article


146. Directional Kernel Mean Difference: A Fast Signed Statistic for Univariate Distribution Comparison ​

Author: Shijie Zhong, Jiangfeng Fu
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.20119v1 Announce Type: cross Abstract: We introduce the Directional Kernel Mean Difference (DKMD), a signed statistic for univariate distribution comparison that preserves the direction of distributional shifts. Unlike the squared Maximum Mean Discrepancy (MMD), which discards directional...

📖 Read original article


147. HeadCast: Casting Attention Heads for Efficient Autoregressive Video Generation ​

Author: Jinliang Shen, Lianghao Su, Zheming Li, Kang He, ZiLiang Lai, Yanbing Jiang, Chengru Song
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.20125v1 Announce Type: cross Abstract: Autoregressive (AR) video diffusion models have become a promising paradigm for long and streaming video synthesis, but the continuously growing Key-Value (KV) cache makes attention the dominant inference cost, especially at high resolution where eac...

📖 Read original article


148. Multi-stage Dynamic Selection for Cross-Project Defect Prediction ​

Author: Juscimara G. Avelino, Juscelino S. A. Junior, George D. C. Cavalcanti, Rafael M. O. Cruz
Published: 7/23/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2607.20151v2 Announce Type: cross Abstract: Cross-Project Defect Prediction (CPDP) involves building models using data from external projects, called training projects, to predict modules from the target project. However, traditional CPDP methods suffer from the distribution shift between trai...

📖 Read original article


149. Plausibility-Driven Prioritization of Candidate Biomedical Annotations ​

Author: Emanuele Cavalleri, Miad Alavinezhad, Dario Malchiodi, Marco Mesiti
Published: 7/23/2026, 4:00:00 AM
Categories: q-bio.QM, cs.DB, cs.LG

arXiv:2607.20163v1 Announce Type: cross Abstract: The rapid growth of biomedical knowledge has made the validation of automatically generated biological annotations a major bottleneck in biomedical curation. While computational methods can rapidly produce large numbers of candidate annotations, dete...

📖 Read original article


150. Hard Guarantees at a Measured Price: Entropy-Stable Learned Finite Volumes for Compressible Flow ​

Author: Denis Gueyffier (ONERA -- Institut Polytechnique de Paris)
Published: 7/23/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG, cs.NA, math.NA

arXiv:2607.20171v1 Announce Type: cross Abstract: Learned solvers for compressible flow are usually compared to classical methods at equal mesh resolution rather than at equal computational cost, and they typically offer no guarantee that their solutions remain physically admissible. We present a le...

📖 Read original article


151. Statistical Inference for Rank Allocation in Low-Rank Adaptation ​

Author: Yihang Gao, Vincent Y. F. Tan
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2607.20205v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has become a widely used parameter-efficient fine-tuning method for large language models. Since different modules and layers may contribute unequally to downstream adaptation, allocating rank resources under a fixed parame...

📖 Read original article


152. Dynamical and Optimization Trade-offs of Levi--Civita Coordinates for Learned Close-Encounter Dynamics ​

Author: Abhishek Shankar
Published: 7/23/2026, 4:00:00 AM
Categories: physics.comp-ph, astro-ph.EP, cs.LG

arXiv:2607.20235v1 Announce Type: cross Abstract: Classical regularization removes the binary-collision singularity from the Kepler problem, but its value as a representation for learned Hamiltonian dynamics has not been systematically isolated. We compare Cartesian and planar Levi--Civita formulati...

📖 Read original article


153. Adaptive Bayesian Online Learning via Expert Aggregation ​

Author: Jungbin Jun, Ilsang Ohn
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.20239v1 Announce Type: cross Abstract: Bayesian online learning promises uncertainty-aware prediction on data streams, but its performance hinges on inferential choices, including learning rates, prior distributions and variational families, which are usually fixed before seeing the strea...

📖 Read original article


154. Self-supervision drives representational convergence in medical foundation models more than clinical supervision ​

Author: Soroosh Tayebi Arasteh, Sebastian Ziegelmayer, Mahshad Lotfinia, Lisa Adams, Sven Nebelung, Jakob Nikolas Kather, Daniel Truhn
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG

arXiv:2607.20274v1 Announce Type: cross Abstract: Medical image encoders from different groups are increasingly treated as interchangeable, on the assumption that scale and clinical supervision concentrate their representations onto a shared structure. Whether this convergence is real, what produces...

📖 Read original article


155. Adaptive deep nonparametric regression from dependent data under covariate shift ​

Author: William Kengne, Ehud Mossa Ockegna
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.20309v1 Announce Type: cross Abstract: Covariate shift often occurs because, in many real applications, the source and the target observations may be generated from different distributions. In this case, the standard metric under the source distribution is not appropriate. This paper cons...

📖 Read original article


156. Decentralized Online Riemannian Optimization for Strongly Geodesically Convex Functions ​

Author: Zhanyuan Cai, Emre Sahinoglu, Shahin Shahrampour
Published: 7/23/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.MA

arXiv:2607.20316v1 Announce Type: cross Abstract: We study decentralized online optimization for strongly geodesically convex (strongly g-convex) losses on Riemannian manifolds with bounded sectional curvature, including positively curved manifolds. In centralized Riemannian optimization, strong g-c...

📖 Read original article


157. Label-Free Finite-Volume-Residual Training of Attention Graph Neural Networks for Coupled Thermo-Fluid Fields ​

Author: Tianyu Li, Zhiwei Cao, Qingang Zhang, Ruihang Wang, Binyang Song, Yonggang Wen
Published: 7/23/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG

arXiv:2607.20321v2 Announce Type: cross Abstract: Neural surrogates are widely used in scientific machine learning for fast prediction of three-dimensional (3D) thermo-fluid fields. However, generating training data using conventional numerical solvers often incurs substantial computational and stor...

📖 Read original article


158. Statevector-Referenced Geometry Survival of a Four-Qubit ZZ Quantum Kernel on IBM Quantum Hardware: A Fixed-Subset Diagnostic Across Three Execution Configurations ​

Author: Rostyslav Sipakov
Published: 7/23/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.20377v1 Announce Type: cross Abstract: Quantum-kernel methods encode a dataset's geometry in a Gram matrix, so learning claims on hardware kernels assume the intended geometry survives execution. We measure that survival for one frozen four-qubit ZZ feature-map kernel on $N=24$ real indoo...

📖 Read original article


159. Towards Miniature Humanoid Tele-Loco-Manipulation Using Virtual Reality and Reinforcement Learning ​

Author: Nicolas Kosanovic, Jordan Dowdy, Jean Chagas Vaz
Published: 7/23/2026, 4:00:00 AM
Categories: cs.RO, cs.HC, cs.LG

arXiv:2607.20399v1 Announce Type: cross Abstract: Full-sized humanoid robot capabilities have grown exponentially in recent years, aiming towards general-purpose deployment in human environments. A popular control method used by manufacturers utilizes Virtual Reality for upper-body teleoperation and...

📖 Read original article


160. Lipschitzian SLLNs for random functions ​

Author: Lai Tian, Johannes O. Royset
Published: 7/23/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.ST, stat.TH

arXiv:2607.20411v1 Announce Type: cross Abstract: We prove strong laws of large numbers for locally Lipschitz functions in the Lipschitz pseudometric. Our results hold under either a topological or a model-theoretic condition, with the latter encompassing functions jointly definable in o-minimal str...

📖 Read original article


161. A convergence result of a continuous model of deep learning via a \L{}ojasiewicz--Simon inequality ​

Author: Noboru Isobe
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, math.AP, math.FA, math.PR

arXiv:2311.15365v3 Announce Type: replace Abstract: We study an idealized training process for deep neural networks in a continuous-depth, mean-field model in which each layer is parameterized by a probability measure on a Euclidean parameter space. The training dynamics are formulated as a Wasserst...

📖 Read original article


162. Differentially Private Neural Network Training Under the Hidden State Assumption ​

Author: Ding Chen, Haochen Luo, Xiaofei Wang, Chen Liu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2407.08233v3 Announce Type: replace Abstract: Current differentially private learning paradigms face a severe utility bottleneck: DP-SGD degrades performance through noise accumulation over training steps, while aggregation-based approaches such as PATE suffer from data inefficiency due to dis...

📖 Read original article


163. Theory-to-Practice Gap for Neural Networks and Neural Operators ​

Author: Philipp Grohs, Samuel Lanthaler, Margaret Trautner
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, math.FA

arXiv:2503.18219v2 Announce Type: replace Abstract: This work studies the sampling complexity of learning with ReLU neural networks and neural operators. For mappings belonging to relevant approximation spaces, we derive upper bounds on the best-possible convergence rate of any learning algorithm, w...

📖 Read original article


164. Interpretable Deep Learning Paradigm for Airborne Transient Electromagnetic Inversion ​

Author: Shuang Wang, Xuben Wang, Fei Deng, Peifan Jiang, Lifeng Mao
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2503.22214v2 Announce Type: replace Abstract: The extraction of geoelectric structural information from airborne transient electromagnetic (ATEM) data primarily involves data processing and inversion. Conventional methods rely on empirical parameter selection, making it difficult to process co...

📖 Read original article


165. DREMnet: An Interpretable Denoising Framework for Semi-Airborne Transient Electromagnetic Signal ​

Author: Shuang Wang, Ming Guo, Xuben Wang, Fei Deng, Lifeng Mao, Bin Wang, Wenlong Gao
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2503.22223v2 Announce Type: replace Abstract: The semi-airborne transient electromagnetic method (SATEM) is capable of conducting rapid surveys over large-scale and hard-to-reach areas. However, the acquired signals are often contaminated by complex noise, which can compromise the accuracy of ...

📖 Read original article


166. AuditVotes: Elevating Provable Defense for GNNs with Efficient Augmentation and Conditional Smoothing ​

Author: Yuni Lai, Yulin Zhu, Yixuan Sun, Yulun Wu, Bin Xiao, Gaolei Li, Jianhua Li, Qi Xie, Kai Zhou
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2503.22998v3 Announce Type: replace Abstract: Despite advancements in Graph Neural Networks (GNNs), adaptive attacks continue to challenge their robustness. Certified robustness via randomized smoothing offers provable guarantees but suffers from a severe accuracy-robustness trade-off, limitin...

📖 Read original article


167. Streaming Sliced Optimal Transport ​

Author: Khai Nguyen
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, stat.CO, stat.ME, stat.ML

arXiv:2505.06835v5 Announce Type: replace Abstract: Sliced optimal transport (SOT), or sliced Wasserstein (SW) distance, is widely recognized for its statistical and computational scalability. In this work, we further enhance computational scalability by proposing the first method for estimating SW ...

📖 Read original article


168. Membership Inference Attacks for Unseen Classes ​

Author: Pratiksha Thaker, Neil Kale, Zhiwei Steven Wu, Virginia Smith
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, stat.ML

arXiv:2506.06488v3 Announce Type: replace Abstract: A key tool in developing safe AI models is \emph{data auditing}, i.e., using statistical tools to determine whether harmful content may have been used in the training data of a black-box model. Unfortunately, most \emph{membership inference attacks...

📖 Read original article


169. OrbitAll: A Unified Quantum Mechanical Representation Deep Learning Framework for All Molecular Systems ​

Author: Beom Seok Kang, Vignesh C. Bhethanabotla, Amin Tavakoli, Maurice D. Hanisch, Arimitsu Horikawa-Strakovsky, Miguel Nouman, Danish Khan, William A. Goddard III, Anima Anandkumar
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph

arXiv:2507.03853v2 Announce Type: replace Abstract: We introduce OrbitAll, a geometry- and physics-informed deep learning framework that encodes any molecular system with arbitrary charges, spins, and environmental effects using electronic structure information. It utilizes spin-polarized orbital fe...

📖 Read original article


170. NeuCoReClass AD: Redefining Self-Supervised Time Series Anomaly Detection ​

Author: Aitor S'anchez-Ferrera, Usue Mori, Borja Calvo, Jose A. Lozano
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2508.00909v2 Announce Type: replace Abstract: Time series anomaly detection plays a critical role in a wide range of real-world applications. Among unsupervised approaches, self-supervised learning has gained traction for modeling normal behavior without the need of labeled data. However, many...

📖 Read original article


171. Label-Noise Resistant Learning via Optimal Brain Damage Masking ​

Author: Xinlei Zhang, Fan Liu, Chuanyi Zhang, Xiaoying Ji, Wenhui Wang, Wei Zhou, Yuhui Zheng
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2508.09697v4 Announce Type: replace Abstract: Noisy labels are inevitable in real-world multimedia applications. Due to the strong memorization capacity of deep neural networks, these noisy labels cause significant performance degradation. Existing noise-robust methods have mainly focused on r...

📖 Read original article


172. Self-Explaining Reinforcement Learning for Mobile Network Resource Allocation ​

Author: Konrad Nowosadko, Franco Ruggeri, Ahmad Terra
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.NI

arXiv:2509.14925v2 Announce Type: replace Abstract: Deep reinforcement learning (DRL) methods, though powerful, often lack transparency, which limits their adoption in critical domains. We apply Self-Explaining Neural Networks (SENNs) to RL by parametrizing the policy of a PPO agent with a SENN, pro...

📖 Read original article


173. On the Separability of Information in Diffusion Models ​

Author: Akhil Premkumar
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, cs.AI, cs.IT, math.IT

arXiv:2509.23937v5 Announce Type: replace Abstract: Diffusion models transform noise into data by injecting information that was captured in their neural network during the training phase. In this paper, we ask: \textit{what} is this information? We find that, in pixel-space diffusion models, (1) a ...

📖 Read original article


174. Failure Modes of Always-On Inter-Cluster Repulsion in Replay-Based Continual Learning ​

Author: Md Hasibul Amin, Tamzid Tanvi Alam
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.07648v3 Announce Type: replace Abstract: Feature-space objectives are often added to replay-based continual learning systems with the expectation that better geometric separation will improve retention. We study a preliminary form of Cluster-Aware Replay (CAR) that combines a class-balanc...

📖 Read original article


175. One4Many-StablePacker: An Efficient Deep Reinforcement Learning Framework for the 3D Bin Packing Problem ​

Author: Lei Gao, Shihong Huang, Shengjie Wang, Hong Ma, Feng Zhang, Hengda Bao, Qichang Chen, Weihua Zhou
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.10057v2 Announce Type: replace Abstract: The three-dimensional bin packing problem (3D-BPP) is widely applied in logistics and warehousing. Existing learning-based approaches often neglect practical stability-related constraints and exhibit limitations in generalizing across diverse bin d...

📖 Read original article


176. Dominant vs. Dominated: Concept-Level Generative Collapse in Diffusion Models ​

Author: Hayeon Jeong, Jong-Seok Lee
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2512.20666v2 Announce Type: replace Abstract: Text-to-image diffusion models have attracted significant attention for their ability to generate diverse, high-fidelity images. However, in multi-concept generation, one concept token often dominates the output while others are suppressed-a phenom...

📖 Read original article


177. Revisiting Forest Proximities via Sparse Leaf-Incidence Kernels ​

Author: Adrien Aumon, Guy Wolf, Kevin R. Moon, Jake S. Rhodes
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, cs.PF

arXiv:2601.02735v3 Announce Type: replace Abstract: Decision forests induce supervised similarities through the partition structure of their trees. Yet forest proximity computation is still often treated as a quadratic operation in the number of samples, which limits scalability and restricts broade...

📖 Read original article


178. ArenaRL: Scaling RL for Open-Ended Agents via Tournament-based Relative Ranking ​

Author: Qiang Zhang, Boli Chen, Fanrui Zhang, Ruixue Ding, Shihang Wang, Qiuchen Wang, Yinfeng Huang, Haonan Zhang, Rongxiang Zhu, Pengyong Wang, Ailin Ren, Xin Li, Pengjun Xie, Jiawei Liu, Ning Guo, Jingren Zhou, Zheng-Jun Zha
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.06487v3 Announce Type: replace Abstract: Reinforcement learning has substantially improved the performance of LLM agents on tasks with verifiable outcomes, but it still struggles on open-ended agent tasks with vast solution spaces (e.g., complex travel planning). Due to the absence of obj...

📖 Read original article


179. Geometric Attention: A Regime-Explicit Operator Semantics for Transformer Attention ​

Author: Luis Rosario Freytes
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.11618v2 Announce Type: replace Abstract: Geometric Attention (GA) specifies an attention layer by four independent inputs: a finite carrier (what indices are addressable), an evidence-kernel rule (how masked proto-scores and a link induce nonnegative weights), a probe family (which observ...

📖 Read original article


180. A Sheaf-Theoretic and Topological Perspective on Complex Network Modeling and Attention Mechanisms in Graph Neural Models ​

Author: Chuan-Shen Hu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.AT

arXiv:2601.21207v4 Announce Type: replace Abstract: Combinatorial and topological structures, such as graphs, simplicial complexes, and cell complexes, form the foundation of geometric and topological deep learning (GDL and TDL) architectures. These models aggregate signals over such domains, integr...

📖 Read original article


181. Generative AI-enhanced Probabilistic Multi-Fidelity Surrogate Modeling Via Transfer Learning ​

Author: Jice Zeng, David Barajas-Solano, Hui Chen
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2602.00072v2 Announce Type: replace Abstract: The performance of machine learning surrogates is critically dependent on data quality and quantity. This presents a major challenge, as high-fidelity (HF) data is often scarce and computationally expensive to acquire, while low-fidelity (LF) data ...

📖 Read original article


182. In-Run Data Shapley for Adam Optimizer ​

Author: Meng Ding, Zeqing Zhang, Di Wang, Lijie Hu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.00329v4 Announce Type: replace Abstract: Reliable data attribution is essential for mitigating bias and reducing computational waste in modern machine learning, with the Shapley value serving as the theoretical gold standard. While recent "In-Run" methods bypass the prohibitive cost of re...

📖 Read original article


183. SwiftRepertoire: Few-Shot Immune-Signature Synthesis via Dynamic Kernel Codes ​

Author: Rong Fu, Muge Qi, Yang Li, Yabin Jin, Jiekai Wu, Chunlei Meng, Juntao Gao, Li Bao, Qi Zhao, Wei Luo, Youjin Wang, Simon Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.01051v5 Announce Type: replace Abstract: Repertoire-level analysis of T cell receptors offers a biologically grounded signal for disease detection and immune monitoring, yet practical deployment is impeded by label sparsity, cohort heterogeneity, and the computational burden of adapting l...

📖 Read original article


184. The Geometry of Learning to Avoid Interventions ​

Author: Ethan Pronovost, Khimya Khetarpal, Siddhartha Srinivasa
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.03825v2 Announce Type: replace Abstract: Human interventions are a common source of supervision in autonomous systems during deployment. Many existing approaches are based on avoiding interventions, yet the consequences of this objective are not well understood. We develop a geometric per...

📖 Read original article


185. NeuroPareto: Calibrated Acquisition for Costly Many-Goal Search in Vast Parameter Spaces ​

Author: Rong Fu, Chunlei Meng, Haoyu Zhao, Kun Liu, JiaBao Dou, Youjin Wang, Simon James Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2602.03901v5 Announce Type: replace Abstract: The pursuit of optimal trade-offs in high-dimensional search spaces under stringent computational constraints poses a fundamental challenge for contemporary multi-objective optimization. We develop NeuroPareto, a cohesive architecture that integrat...

📖 Read original article


186. AdvSynGNN: Structure-Adaptive Graph Neural Nets via Adversarial Synthesis and Self-Corrective Propagation ​

Author: Rong Fu, Muge Qi, Chunlei Meng, Shuo Yin, Kun Liu, Simon Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17071v3 Announce Type: replace Abstract: Graph neural networks frequently encounter significant performance degradation when confronted with structural noise or non-homophilous topologies. To address these systemic vulnerabilities, we present AdvSynGNN, a comprehensive architecture design...

📖 Read original article


187. SubQuad: Near-Quadratic-Free Structure Inference with Distribution-Balanced Objectives in Adaptive Receptor framework ​

Author: Rong Fu, Zijian Zhang, Kun Liu, Jiekai Wu, Xianda Li, Simon Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17330v5 Announce Type: replace Abstract: Comparative analysis of adaptive immune repertoires at population scale is hampered by two practical bottlenecks: the near-quadratic cost of pairwise affinity evaluations and dataset imbalances that obscure clinically important minority clonotypes....

📖 Read original article


188. TempoNet: Slack-Quantized Transformer-Guided Reinforcement Scheduler for Adaptive Deadline-Centric Real-Time Dispatchs ​

Author: Rong Fu, Yibo Meng, Zeyu Zhang, Ziming Guo, Jia Yee Tan, Xiaojing Du, Simon James Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.OS, cs.SY, eess.SY

arXiv:2602.18109v3 Announce Type: replace Abstract: Real-time schedulers must reason about tight deadlines under strict compute budgets. We present TempoNet, a reinforcement learning scheduler that pairs a permutation-invariant Transformer with a deep Q-approximation. An Urgency Tokenizer discretize...

📖 Read original article


189. Frequentist Consistency of Prior-Data Fitted Networks for Causal Inference ​

Author: Valentyn Melnychuk, Vahid Balazadeh, Stefan Feuerriegel, Rahul G. Krishnan
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.12037v3 Announce Type: replace Abstract: Foundation models based on prior-data fitted networks (PFNs) have shown strong empirical performance in causal inference by framing the task as an in-context learning problem. However, it is unclear whether PFN-based causal estimators provide uncer...

📖 Read original article


190. Pre-Deployment Complexity Estimation for Federated Perception Systems ​

Author: KMA Solaiman, Shafkat Islam, Ruy de Oliveira, Bharat Bhargava
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC

arXiv:2603.28282v2 Announce Type: replace Abstract: Edge AI systems increasingly rely on federated learning to train perception models in distributed, privacy-preserving, and resource-constrained environments. Before training, however, practitioners often lack practical tools for estimating task dif...

📖 Read original article


191. A Unified Survival Benchmark for Temporal Dropout Risk Prediction in Learning Analytics ​

Author: Rafael da Silva, Jeff Eicher, Gregory Longo
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.08870v3 Announce Type: replace Abstract: Student dropout is a persistent concern in Learning Analytics, yet comparative studies frequently evaluate predictive models under heterogeneous protocols, prioritizing discrimination over temporal interpretability and calibration. This study intro...

📖 Read original article


192. An Auditable Policy-Simulation Framework for Student Dropout in Intervention-Free Data ​

Author: Rafael da Silva, Jeff Eicher, Gregory Longo
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.08874v3 Announce Type: replace Abstract: This study proposes a temporal modeling framework with a counterfactual policy-simulation layer for student dropout in higher education, using LMS engagement data and administrative withdrawal records. Dropout is operationalized as a time-to-event ...

📖 Read original article


193. Material-agnostic temperature field prediction for metal additive manufacturing via a parametric PINN framework ​

Author: Hyeonsu Lee, Jihoon Jeong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, physics.app-ph, physics.comp-ph

arXiv:2604.14562v2 Announce Type: replace Abstract: Accurate temperature field prediction in metal additive manufacturing (AM) is essential for understanding the process-structure-performance relationship. While prior studies have explored generalization to unseen process conditions, they often requ...

📖 Read original article


194. Generative Augmented Inference of LLM-generated Data for Market Research: Theory and Empirical Evidence ​

Author: Cheng Lu, Mengxin Wang, Dennis J. Zhang, Heng Zhang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ME, stat.ML

arXiv:2604.14575v3 Announce Type: replace Abstract: Marketing research often relies on parameters estimated from costly human-generated data, such as conjoint survey responses, purchase decisions, and field experiment outcomes. Recent advances in large language models (LLMs) and other AI systems off...

📖 Read original article


195. OC-Distill: Ontology-aware Contrastive Learning with Cross-Modal Distillation for ICU Risk Prediction ​

Author: Zhongyuan Liang, Junhyung Jo, Hyang-Jung Lee, Sang Kyu Kim, Irene Y. Chen
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.16878v3 Announce Type: replace Abstract: Early prediction of severe clinical deterioration and remaining length of stay can enable timely intervention and better resource allocation in high-acuity settings such as the ICU. This has driven the development of machine learning models that le...

📖 Read original article


196. FES-FM: Free Energy Surface Sampling via Reduced Flow Matching ​

Author: Zichen Liu, Tiejun Li
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.00337v2 Announce Type: replace Abstract: Sampling the distribution of collective variables (CVs) and estimating the associated free energy surface are crucial problems in statistical physics, as they underpin a better understanding of chemical reactions and conformational transitions. Tra...

📖 Read original article


197. Tunable MAGMAX: Preference-Aware Model Merging for Continual Learning ​

Author: Kei Hiroshima, Kento Uchida, Shinichi Shirakawa
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.20803v3 Announce Type: replace Abstract: Continual learning (CL) aims to train models sequentially on multiple tasks while mitigating catastrophic forgetting of previously learned knowledge. Recent advances in large pre-trained models (LPMs) and model merging techniques, such as MAGMAX, h...

📖 Read original article


198. A Geometric Approach to Constrained Online Learning ​

Author: Dhruv Sarkar, Abhishek Sinha
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2605.21107v3 Announce Type: replace Abstract: We study constrained online convex optimization with adversarial time-varying constraints. At each round the learner acts before observing the loss and constraint, and is compared with the best fixed action satisfying all constraints in hindsight. ...

📖 Read original article


199. Reading Calibrated Uncertainty from Language Model Trajectories ​

Author: Aliai Eusebi, Alexander Herzog, Xiaoyu Liang, Marie Vasek, Enrico Mariconti, Lorenzo Cavallaro
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.22864v2 Announce Type: replace Abstract: The maximum softmax probability (MSP) represents a default approach when evaluating uncertainty quantification for language model generation with structured output. Although cheap, it is often miscalibrated. Methods that probe the model's internal ...

📖 Read original article


200. A transition-density-based operator learning method for Fokker-Planck equations with various initial conditions ​

Author: Li Zeng, Xiaoliang Wan, Yaobin Wang, Fabio Nobile, Tao Zhou
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.09434v2 Announce Type: replace Abstract: Solving Fokker-Planck equations (FPEs) for multiple initial conditions typically requires repeated computations, leading to substantial computational costs. In this work, we propose a transition-density-based operator learning method to efficiently...

📖 Read original article


201. ReRAM-aware Model Finetuning addressing I-V Non-linearity and Retention Errors ​

Author: Ching-Yi Lin, Shamik Kundu, Arnab Raha, Sahil Shah
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2606.17471v3 Announce Type: replace Abstract: Traditional CPU, GPU, and NPU architectures are increasingly limited by the von Neumann bottleneck. While In-Memory Computing (IMC) using ReRAM crossbar arrays offers a high-density, energy-efficient alternative, its practical deployment is constra...

📖 Read original article


202. Reducing Learner Redundancy in Boosting via Residual Orthogonalization ​

Author: Ye Su, Jipeng Guo, Xin Xu, Gangchun Zhang, Jinxin Chen, Di Wu, Longlong Zhao
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.17567v2 Announce Type: replace Abstract: While sequential residual fitting is the bedrock of standard boosting frameworks, it inherently breeds learner redundancy by repeatedly revisiting correlated error components. To address this bottleneck, we propose a shift from residual fitting to ...

📖 Read original article


203. Spectral DPPs via NEPv: A Scalable Continuous Relaxation of Determinantal MAP for Diversity-Aware Data Selection ​

Author: Richard Yi Da Xu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.19411v2 Announce Type: replace Abstract: Selecting a small, diverse, high-quality subset from a massive pool of candidates is a recurring primitive in modern machine learning -- data curation and coreset selection for training and fine-tuning large models, active-learning batch acquisitio...

📖 Read original article


204. Boundary Embedding Shaping with Adaptive Contrastive Learning for Graph Structural Disentanglement ​

Author: Jiaqing Chen, Zidu Yin, Yichao Cai, Yuhang Liu, Zhen Zhang, Dong Gong, Javen Qinfeng Shi
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.20283v2 Announce Type: replace Abstract: Graph neural networks (GNNs) excel at aggregating neighbor information for classification, yet their performance is hindered by graph structural entanglement, where spurious correlations from semantically irrelevant neighbors contaminate node embed...

📖 Read original article


205. Safety-Regulated Transfer Reinforcement Learning with Adaptive Teacher Guidance ​

Author: Wenjie Huang, Yang Li, Jingjia Teng, Mingwei Jin, Kai Song, Zeyu Yang, Qisong Yang, Yougang Bian
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.26527v2 Announce Type: replace Abstract: We propose Safety-Regulated Adaptive Transfer Reinforcement Learning (SRATRL), a teacher--student framework that combines safety-triggered intervention, safety-adaptive value shaping, and policy-compatibility-based optimization for efficient target...

📖 Read original article


206. Sketched Linear Contrastive Learning: Approximation, Optimization, and Statistical Scaling ​

Author: Ziyan Chen, Zhongzhu Zhou, Ding-Xuan Zhou
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.26617v2 Announce Type: replace Abstract: Scaling laws describe how learning performance varies with model size, data size, and compute. While recent theoretical work has established scaling laws for sketched linear regression, much less is understood for contrastive representation learnin...

📖 Read original article


207. Experience Augmented Policy Optimization for LLM Reasoning ​

Author: Jinda Lu, Kexin Huang, Junkang Wu, Shuo Yang, Jinghan Li, Chiyu Ma, Shaohang Wei, Xiang Wang, Guoyin Wang, Jingren Zhou
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.30420v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a powerful paradigm for improving the reasoning capabilities of large language models (LLMs). However, existing RLVR methods typically rely on on-policy optimization from scratch, resulting i...

📖 Read original article


208. In-span learning: adapting reduced-order models using their own predictions ​

Author: Amirpasha Hedayat, Laura Balzano, Karthik Duraisamy
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.NA, math.NA

arXiv:2607.02937v2 Announce Type: replace Abstract: Reduced-order models compress high-dimensional dynamics into low-dimensional representations that can be evaluated rapidly, but they lose accuracy when online dynamics drift beyond the training data. Adaptive methods address this by updating the su...

📖 Read original article


209. Scaling Time Series Classification via XAI-Driven Data Reduction ​

Author: Davide Italo Serramazza, Thach Le Nguyen, Georgiana Ifrim
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.15774v2 Announce Type: replace Abstract: Explainable AI (XAI) for time series has seen significant algorithmic growth, but its utility in providing measurable performance gains for downstream tasks remains under-explored. This paper bridges this gap by introducing drXAI, a novel methodolo...

📖 Read original article


210. The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models ​

Author: Francesco Karim Vicidomini
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.16741v2 Announce Type: replace Abstract: B"urger et al. (2024) demonstrated that truth representations in large language models are universal across statement polarity but reside within a multidimensional subspace. We extend this framework along three questions: how the dimensionality of...

📖 Read original article


211. JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models ​

Author: Ruiyi Ding, Jie Li, He Kang, Ziyan Liu, Chengru Song, Yuan chen
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.SY, eess.SY

arXiv:2607.17572v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) is a powerful reinforcement learning algorithm for aligning generative models with human preferences. While successful in large language models~\cite{shao2024deepseekmathpushinglimitsmathematical}, its exte...

📖 Read original article


212. ChemHyperMag: Physics-informed magnetic hypergraph learning improves molecular ADMET prediction ​

Author: Hexiao Ding, Hongzhao Chen, Jing Lan, Yufeng Jiang, Zihong Luo, Zehua Xiong, Tianlong Ruan, Yunlin Mao, Nga Chun Ng, Gwing Kei Yip, Gerald W. Y. Cheng, Kate Inyoung Oh, Jing Cai, Liang-Ting Lin, Jung Sun Yoo
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.18332v2 Announce Type: replace Abstract: Accurate prediction of ADMET (Absorption, Distribution, Metabolism, Excretion, and Toxicity) is important for drug discovery. Most predictors use undirected molecular graphs and pairwise edges. This choice misses asymmetric interactions, nonreversi...

📖 Read original article


213. Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning ​

Author: Junyao Yang, Yucheng Shi, Zongxia Li, Zhongzhi Li, Ruhan Wang, Xiangxin Zhou, Kishan Panaganti, Haitao Mi, Leowei Liang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.18722v2 Announce Type: replace Abstract: Asynchronous reinforcement learning improves throughput by decoupling rollout generation from optimization, but staleness is an inevitable byproduct compounded by policy lag, engine delays, and mixture-of-experts routing. From a trust-region perspe...

📖 Read original article


214. H$^2$SD: Hybrid Hindsight Self-Distillation ​

Author: Qiye Cai, Yichuan Ma, Linyang Li, Peiji Li, Yongkang Chen, Qipeng Guo, Yicheng Zou, Xiaocheng Feng, Bing Qin
Published: 7/23/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.18955v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) provides reliable outcome supervision for language model reasoning, but a scalar trajectory reward offers limited token-level guidance. Existing self-distillation methods add a privileged teache...

📖 Read original article


215. Auto-adaptive Resonance Equalization using Dilated Residual Networks ​

Author: Maarten Grachten, Emmanuel Deruty, Alexandre Tanguy
Published: 7/23/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS

arXiv:1807.08636v2 Announce Type: replace-cross Abstract: In music and audio production, attenuation of spectral resonances is an important step towards a technically correct result. In this paper we present a two-component system to automate the task of resonance equalization. The first component i...

📖 Read original article


216. Kernel Ridge Regression Inference ​

Author: Rahul Singh, Suhas Vijaykumar
Published: 7/23/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.ML, stat.TH

arXiv:2302.06578v4 Announce Type: replace-cross Abstract: We provide uniform confidence bands for kernel ridge regression (KRR), a widely used nonparametric regression estimator for nonstandard data such as preferences, sequences, and graphs. Despite the prevalence of these data--e.g., student prefe...

📖 Read original article


217. Any-Time Regret-Guaranteed Algorithm for Control of Linear Quadratic Systems ​

Author: Jafar Abbaszadeh Chekan, Cedric Langbort
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.SY, eess.SY

arXiv:2406.07746v4 Announce Type: replace-cross Abstract: We propose a computationally efficient algorithm that achieves anytime regret of order $\mathcal{O}(\sqrt{t})$, with explicit dependence on the system dimensions and on the solution of the Discrete Algebraic Riccati Equation (DARE). Our appro...

📖 Read original article


218. LaSEr-Edit: Localized Span-level Error Editing with Energy-based Localization ​

Author: Hye Ryung Son, Saehee Eom, Mooho Song, Jay-Yoon Lee
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2407.00740v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are widely adopted in real-world applications, it has become critical to ensure LLMs satisfy safety constraints, such as non-toxicity and logical consistency, as well as task- and situation-specific constraints...

📖 Read original article


219. A Confidence Interval for the $\ell_2$ Expected Calibration Error ​

Author: Yan Sun, Pratik Chaudhari, Ian J. Barnett, Edgar Dobriban
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2408.08998v4 Announce Type: replace-cross Abstract: Recent advances in machine learning have significantly improved prediction accuracy in various applications. However, ensuring the calibration of probabilistic predictions remains a significant challenge. Despite efforts to enhance model cali...

📖 Read original article


220. Distributed Optimization via Energy Conservation Laws in Dilated Coordinates ​

Author: Kushal Chakrabarti, Mayank Baranwal
Published: 7/23/2026, 4:00:00 AM
Categories: math.OC, cs.AI, cs.LG, cs.SY, eess.SY, math.DS

arXiv:2409.19279v2 Announce Type: replace-cross Abstract: Continuous-time models can reveal accelerated structures in distributed optimization, but their rates need not survive direct discretization. We introduce a second-order primal--dual flow for smooth convex distributed optimization and constru...

📖 Read original article


221. SEED: Towards More Accurate Semantic Evaluation for Visual Brain Decoding ​

Author: Juhyeon Park, Peter Yongho Kim, Jiook Cha, Shinjae Yoo, Taesup Moon
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2503.06437v3 Announce Type: replace-cross Abstract: We present SEED (Semantic Evaluation for Visual Brain Decoding), a novel metric for evaluating the semantic decoding performance of visual brain decoding models. It integrates three complementary metrics, each capturing a different aspect of ...

📖 Read original article


222. Towards Practical Emotion Recognition: An Unsupervised Source-Free Approach for EEG Domain Adaptation ​

Author: Md Niaz Imtiaz, Naimul Khan
Published: 7/23/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2504.03707v2 Announce Type: replace-cross Abstract: Emotion recognition is crucial for advancing mental health, healthcare, and technologies such as brain-computer interfaces. EEG-based models, however, struggle in cross-domain settings due to the high cost of labeled data and signal variabili...

📖 Read original article


223. A Novel Hybrid Deep Learning Technique for Speech Emotion Detection using Feature Engineering ​

Author: Shahana Yasmin Chowdhury, Bithi Banik, Md Tamjidul Hoque, Shreya Banerjee
Published: 7/23/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG, eess.AS

arXiv:2507.07046v3 Announce Type: replace-cross Abstract: Nowadays, speech emotion recognition (SER) plays a vital role in the field of human-computer interaction (HCI) and the evolution of artificial intelligence (AI). Our proposed DCRF-BiLSTM model is used to recognize seven emotions: neutral, hap...

📖 Read original article


224. Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review ​

Author: Siyi Hu, Mohamad A Hady, Jianglin Qiao, Jimmy Cao, Mahardhika Pratama, Ryszard Kowalczyk
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2507.10142v2 Announce Type: replace-cross Abstract: Multi-Agent Reinforcement Learning (MARL) has achieved strong performance in simulated benchmarks, yet real deployments often violate the assumptions under which algorithms are designed and evaluated. Agent populations may change, objectives ...

📖 Read original article


225. Comparing Model-agnostic Feature Selection Methods through Relative Efficiency ​

Author: Chenghui Zheng, Garvesh Raskutti
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2508.14268v2 Announce Type: replace-cross Abstract: Feature selection and importance estimation in a model-agnostic setting is an ongoing challenge of significant interest. Wrapper methods are commonly used because they are typically model-agnostic. In this paper, we develop a general comparis...

📖 Read original article


226. In-the-Flow Agentic System Optimization for Effective Planning and Tool Use ​

Author: Zhuofeng Li, Haoxiang Zhang, Seungju Han, Sheng Liu, Jianwen Xie, Yu Zhang, Yejin Choi, James Zou, Pan Lu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2510.05592v2 Announce Type: replace-cross Abstract: Outcome-driven reinforcement learning has advanced reasoning in large language models (LLMs), but prevailing tool-augmented approaches train a single, monolithic policy that interleaves thoughts and tool calls under full context; this scales ...

📖 Read original article


227. Schr\"odinger Bridge Mamba for One-Step Speech Enhancement ​

Author: Jing Yang, Sirui Wang, Chao Wu, Lei Guo, Fan Fan
Published: 7/23/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG, eess.AS

arXiv:2510.16834v3 Announce Type: replace-cross Abstract: We present Schr"odinger Bridge Mamba (SBM), a novel model for efficient speech enhancement by integrating the Schr"odinger Bridge (SB) training paradigm and the Mamba architecture. Experiments of joint denoising and dereverberation tasks de...

📖 Read original article


228. PGTT: Phase-Guided Terrain Traversal for Perceptive Legged Locomotion ​

Author: Alexandros Ntagkas, Chairi Kiourt, Konstantinos Chatzilygeroudis
Published: 7/23/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2510.18348v2 Announce Type: replace-cross Abstract: State-of-the-art perceptive Reinforcement Learning controllers for legged robots typically either (i) impose oscillator-or IK-based gait priors that constrain the action space, bias policy optimization, and limit adaptability across robot mor...

📖 Read original article


229. Geometry-Guided Generative Representation for Functional Brain Graphs ​

Author: Subati Abulikemu, Tiago Azevedo, Michail Mamalakis, John Suckling
Published: 7/23/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG

arXiv:2511.04539v2 Announce Type: replace-cross Abstract: In network neuroscience, functional brain systems are often characterized using separate yet related graph-theoretic or spectral descriptors, overlooking how these properties covary and partially overlap across individuals and conditions. We ...

📖 Read original article


230. WorldPack: Dynamic Frame Compression for Long-context Video World Modeling ​

Author: Yuta Oshima, Yusuke Iwasawa, Masahiro Suzuki, Yutaka Matsuo, Hiroki Furuta
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2512.02473v2 Announce Type: replace-cross Abstract: Video world models have attracted significant attention for their ability to produce high-fidelity future visual observations conditioned on past observations and navigation actions. However, achieving temporally and spatially consistent gene...

📖 Read original article


231. Enhancing next token prediction based pre-training for jet foundation models ​

Author: Joschka Birk, Anna Hallin, Gregor Kasieczka, Nikol Madzharova, Ian Pang, David Shih
Published: 7/23/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-ex, physics.data-an

arXiv:2512.04149v2 Announce Type: replace-cross Abstract: Next token prediction is an attractive pre-training task for jet foundation models, in that it is simulation free and enables excellent generative capabilities that can transfer across datasets. Here we study multiple improvements to next tok...

📖 Read original article


232. Model Gateway: Management Platform for Model-Driven Drug Discovery ​

Author: Yan-Shiun Wu, Sai Mahit Vaddadi, Zachary A. Rollins, Nathan A. Morin
Published: 7/23/2026, 4:00:00 AM
Categories: cs.SE, cs.DC, cs.LG, q-bio.QM

arXiv:2512.05462v2 Announce Type: replace-cross Abstract: Pharmaceutical drug discovery demands machine learning (ML) infrastructure that goes beyond general-purpose Machine Learning Operations (MLOps): inference-time composition of multiple models for multi-parameter optimization (MPO), version man...

📖 Read original article


233. Learning About Learning: A Path from Spin Glasses to Artificial Intelligence ​

Author: Denis D. Caprioti, Matheus Haas, Constantino F. Vasconcelos, Mauricio Girardi-Schappo
Published: 7/23/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cs.AI, cs.LG, physics.comp-ph, physics.ed-ph

arXiv:2601.07635v3 Announce Type: replace-cross Abstract: The Hopfield model, originally inspired by spin glasses, occupies a central place at the intersection of statistical mechanics, neural networks, and artificial intelligence. Despite its conceptual simplicity and broad applicability, it is rar...

📖 Read original article


234. Diffusion-based Annealed Boltzmann Generators : benefits, pitfalls and hopes ​

Author: Louis Grenioux, Maxence Noble
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2601.21026v2 Announce Type: replace-cross Abstract: Sampling configurations at thermodynamic equilibrium is a central challenge in statistical physics. Boltzmann Generators (BGs) tackle it by combining a generative model with a Monte Carlo (MC) correction step to obtain asymptotically unbiased...

📖 Read original article


235. Chimera: Neuro-Symbolic Attention Primitives for Trustworthy Dataplane Intelligence ​

Author: Rong Fu, Xiaowen Ma, Kun Liu, Wangyu Wu, Ziyu Kong, Jia Yee Tan, Tailong Luo, Xianda Li, Yongtai Liu, Youjin Wang, Simon Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.NI, cs.AI, cs.CR, cs.LG

arXiv:2602.12851v4 Announce Type: replace-cross Abstract: Deploying expressive learning models directly on programmable dataplanes promises line-rate, low-latency traffic analysis but remains hindered by strict hardware constraints and the need for predictable, auditable behavior. Chimera introduces...

📖 Read original article


236. Statistical Early Stopping for Reasoning Models ​

Author: Yangxinyu Xie, Tao Wang, Soham Mallick, Yan Sun, Georgy Noarov, Mengxin Yu, Tanwi Mallick, Weijie J. Su, Edgar Dobriban
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, stat.ML

arXiv:2602.13935v2 Announce Type: replace-cross Abstract: While LLMs have seen substantial improvement in reasoning capabilities, they also sometimes overthink, generating unnecessary reasoning steps, particularly under uncertainty, given ill-posed or ambiguous queries. We introduce statistically pr...

📖 Read original article


Author: Rong Fu, Jia Yee Tan, Chunlei Meng, Shuo Yin, Xiaowen Ma, Wangyu Wu, Muge Qi, Simon Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2602.15423v4 Announce Type: replace-cross Abstract: As the burgeoning power requirements of sophisticated neural architectures escalate, the information retrieval community has recognized ecological sustainability as a pivotal priority that necessitates a fundamental paradigm shift in model de...

📖 Read original article


238. Edge-Local and Qubit-Efficient Quantum Graph Learning for the NISQ Era ​

Author: Armin Ahmadkhaniha, Jake Doliskani
Published: 7/23/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG

arXiv:2602.16018v2 Announce Type: replace-cross Abstract: Graph neural networks (GNNs) are a powerful framework for learning representations from graph-structured data, but their direct implementation on near-term quantum hardware remains challenging due to circuit depth, multi-qubit interactions, a...

📖 Read original article


239. Missing-by-Design: Certifiable Modality Deletion for Revocable Multimodal Sentiment Analysis ​

Author: Rong Fu, Ziming Wang, Chunlei Meng, Jiekai Wu, Kangan Qian, Hao Zhang, Simon Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2602.16144v4 Announce Type: replace-cross Abstract: As multimodal systems increasingly process sensitive personal data, the ability to selectively revoke specific data modalities has become a critical requirement for privacy compliance and user autonomy. We present Missing-by-Design (MBD), a u...

📖 Read original article


240. Emotion Collider: Dual Hyperbolic Mirror Manifolds for Sentiment Recovery via Anti Emotion Reflection ​

Author: Rong Fu, Ziming Wang, Shuo Yin, Haiyun Wei, Kun Liu, Xianda Li, Simon Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.MM, cs.CL, cs.LG

arXiv:2602.16161v4 Announce Type: replace-cross Abstract: Emotional expression underpins natural communication and effective human-computer interaction. We present Emotion Collider (EC-Net), a hyperbolic hypergraph framework for multimodal emotion and sentiment modeling. EC-Net represents modality h...

📖 Read original article


241. LiveGraph: Active-Structure Neural Re-ranking for Exercise Recommendation ​

Author: Rong Fu, Zijian Zhang, Haiyun Wei, Jiekai Wu, Kun Liu, Xianda Li, Haoyu Zhao, Yang Li, Yongtai Liu, Ziming Wang, Rui Lu, Simon Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2602.17036v4 Announce Type: replace-cross Abstract: The continuous expansion of digital learning environments has catalyzed the demand for intelligent systems capable of providing personalized educational content. While current exercise recommendation frameworks have made significant strides, ...

📖 Read original article


242. CityGuard: Graph-Aware Private Descriptors for Bias-Resilient Identity Search Across Urban Cameras ​

Author: Rong Fu, Yibo Meng, Jia Yee Tan, Rui Lu, Jiekai Wu, Simon Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2602.18047v4 Announce Type: replace-cross Abstract: City-scale person re-identification across distributed cameras must handle severe appearance changes from viewpoint, occlusion, and domain shift while complying with data protection rules that prevent sharing raw imagery. We introduce CityGua...

📖 Read original article


243. Approximate Nearest Neighbor Search for Modern AI: A Projection-Augmented Graph Approach ​

Author: Kejing Lu, Zhenpeng Pan, Jianbin Qin, Yoshiharu Ishikawa, Chuan Xiao
Published: 7/23/2026, 4:00:00 AM
Categories: cs.IR, cs.DB, cs.LG

arXiv:2603.06660v2 Announce Type: replace-cross Abstract: Approximate Nearest Neighbor Search (ANNS) is fundamental to modern AI applications. Most existing solutions optimize query efficiency but fail to align with the practical requirements of modern workloads. In this paper, we outline six critic...

📖 Read original article


244. Trusting What You Cannot See: Auditable Fine-Tuning and Inference for Proprietary AI ​

Author: Heng Jin, Chaoyu Zhang, Hexuan Yu, Shanghao Shi, Ning Zhang, Y. Thomas Hou, Wenjing Lou
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2603.07466v2 Announce Type: replace-cross Abstract: Cloud-based infrastructure has become the dominant platform for deploying large models, particularly large language models (LLMs). Fine-tuning and inference are increasingly delegated to cloud providers for simplified deployment and access to...

📖 Read original article


245. Posterior Sampling Reinforcement Learning with Gaussian Processes for Continuous Control: Sublinear Regret Bounds for Unbounded State Spaces ​

Author: Hamish Flynn, Joe Watson, Ingmar Posner, Jan Peters
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2603.08287v3 Announce Type: replace-cross Abstract: We analyze the Bayesian regret of the Gaussian process posterior sampling reinforcement learning (GP-PSRL) algorithm. Posterior sampling is a heuristic for decision-making under uncertainty that has been used to develop successful algorithms ...

📖 Read original article


246. IConE: Batch Independent Collapse Prevention for Self-Supervised Representation Learning ​

Author: Konstantinos Almpanakis, Anna Kreshuk
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2603.15263v2 Announce Type: replace-cross Abstract: Self-supervised learning (SSL) has revolutionized representation learning, with Joint-Embedding Architectures (JEAs) emerging as an effective approach for capturing semantic features. Existing JEAs rely on implicit or explicit batch interacti...

📖 Read original article


247. SwiftGS: Episodic Priors for Immediate Satellite Surface Recovery ​

Author: Rong Fu, Jiekai Wu, Haiyun Wei, Xiaowen Ma, Shiyin Lin, Kangan Qian, Chuang Liu, Jianyuan Ni, Simon James Fong
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2603.18634v3 Announce Type: replace-cross Abstract: Rapid, large-scale 3D reconstruction from multi-date satellite imagery is vital for environmental monitoring, urban planning, and disaster response, yet remains difficult due to illumination changes, sensor heterogeneity, and the cost of per-...

📖 Read original article


248. Spectral-transport stability and benign overfitting for minimum norm interpolation ​

Author: Gustav Olaf Yunus Laitinen-Fredriksson Lundstr"om-Imanov
Published: 7/23/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2604.08625v3 Announce Type: replace-cross Abstract: Benign overfitting describes the ability of minimum norm interpolating estimators to generalize despite fitting noisy data exactly. Existing characterizations depend on delicate spectral functionals of the population covariance operator, name...

📖 Read original article


249. Data-Efficient Indentation Size Effect Correction in Steels Using Machine Learning and Physics-Constrained Neural Network ​

Author: Radmir Karamov, Tagir Karamov
Published: 7/23/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG

arXiv:2604.27775v3 Announce Type: replace-cross Abstract: Shallow nanoindentation enables mechanical characterization of thin films, individual phases, and other volume-constrained materials, but the measured hardness is inflated by the indentation size effect (ISE). Classical corrections such as Ni...

📖 Read original article


250. Temporal Fair Division in Multi-Agent Systems: From Precise Alternation Metrics to Scalable Coordination Proxies ​

Author: Nikolaos Al. Papadopoulos, Ismael Tito Freire, Marti Sanchez-Fibla, Konstantinos E. Psannis
Published: 7/23/2026, 4:00:00 AM
Categories: cs.MA, cs.GT, cs.LG

arXiv:2605.14879v2 Announce Type: replace-cross Abstract: Many intelligent computing and autonomous systems rely on multiple independent, often learning, agents repeatedly sharing a limited resource. Examples include autonomous robots accessing a shared workstation, wireless devices competing for co...

📖 Read original article


251. QLIF-CAST: Quantum Leaky-Integrate-and-Fire for Time-Series Weather Forecasting ​

Author: Alberto Marchisio, Aayan Ebrahim, Nouhaila Innan, Muhammad Kashif, Muhammad Shafique
Published: 7/23/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2605.18333v2 Announce Type: replace-cross Abstract: Accurate and efficient time-series forecasting remains a challenging problem for both classical and quantum neural architectures, particularly in multivariate environmental settings. This work adapts the Quantum Leaky Integrate-and-Fire (QLIF...

📖 Read original article


252. TWINGS: Thin Plate Splines Warp-aligned Initialization for Sparse-View Gaussian Splatting ​

Author: Hyeseong Kim, Geonhui Son, Deukhee Lee, Dosik Hwang
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2605.22069v3 Announce Type: replace-cross Abstract: Novel view synthesis from sparse-view inputs poses a significant challenge in 3D computer vision, particularly for achieving high-quality scene reconstructions with limited viewpoints. We introduce TWINGS, a framework that enhances 3D Gaussia...

📖 Read original article


253. Q-PhotoNAS: Hybrid Quantum Neural Architecture Search Framework on Photonic Devices ​

Author: Farah Elnakhal, Alberto Marchisio, Nouhaila Innan, Gabriel Falcao, Muhammad Shafique
Published: 7/23/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2605.22097v2 Announce Type: replace-cross Abstract: Photonic quantum computing is a promising platform for scalable quantum machine learning, but designing effective hybrid architectures remains challenging under hardware and optimization constraints. Existing approaches rely on manually tuned...

📖 Read original article


254. Stability of Low-Rank Implicit Regularization in Perturbed Deep Matrix Factorization ​

Author: Jingzhe Wang, Hung-Hsu Chou
Published: 7/23/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ML

arXiv:2605.28613v2 Announce Type: replace-cross Abstract: This paper studies the stability of low-rank implicit regularization in deep matrix factorization, a tractable model for understanding how gradient-based training can favor low-complexity structure. We first revisit the noiseless setting and ...

📖 Read original article


255. SGMD: Score Gradient Matching Distillation for Few-Step Video Diffusion Distillation ​

Author: Zhuguanyu Wu, Ruihao Gong, Yang Yong, Yushi Huang, Xiangyu Fan, Lei Yang, Dahua Lin, Xianglong Liu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2605.30116v2 Announce Type: replace-cross Abstract: Distribution Matching Distillation (DMD) is a widely used paradigm for accelerating inference in few-step video diffusion models. However, DMD-style video distillation faces two coupled challenges: the fake score must track a continuously evo...

📖 Read original article


256. Don't Fool Me Twice: Adapting to Adversity in the Wild with Experience-Driven Reasoning ​

Author: Navin Sriram Ravie, Andrew Jong, Krrish Jain, John Liu, Omar Alama, Bijo Sebastian, Sebastian Scherer
Published: 7/23/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2605.31119v2 Announce Type: replace-cross Abstract: In robotics, dangers and adversity modes are often embodiment-specific and relative to each agent. A frontier of autonomous mobile robotics is to enable agents to operate effectively in the wild in unseen unstructured environments. A signific...

📖 Read original article


257. Fara-1.5: Scalable Learning Environments for Computer Use Agents ​

Author: Ahmed Awadallah, Sahil Gupta, Yash Lara, Yadong Lu, Hussein Mozannar, Akshay Nambi, Zach Nussbaum, Yash Pandya, Aravind Rajeswaran, Corby Rosset, Alexey Taymanov, Luiz do Valle, Vibhav Vineet, Spencer Whitehead, Andrew Zhao
Published: 7/23/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.20785v2 Announce Type: replace-cross Abstract: Collecting computer use data from human demonstrations is expensive and slow, motivating the need for scalable generation strategies. This requires two key ingredients: environments in which agents can act and verifiers that can judge whether...

📖 Read original article


258. Test-Input Generation for Tensor Programs: What Actually Finds Kernel Bugs ​

Author: Dipankar Sarkar
Published: 7/23/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2606.27396v2 Announce Type: replace-cross Abstract: Test-input generation for tensor kernels is folkloric. Most projects pick a representative shape and dtype, run a fixed-shape allclose-style check, and ship. We make the choices explicit and measure them. Using the gpuemu op-schema-aware seed...

📖 Read original article


259. LoRA-Tuned Large Language Models for Dementia Detection via Multi-View Speech-Derived Features ​

Author: Jonghyeon Park, Olivier Jiyoun Jung, Myungwoo Oh
Published: 7/23/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.CL, cs.LG

arXiv:2606.28445v2 Announce Type: replace-cross Abstract: Early detection of dementia enables timely intervention, and reflecting cognitive impairment, spontaneous speech offers a non-invasive screening modality. Conventional approaches often focus on a single representational dimension -- such as a...

📖 Read original article


260. Format-Controlled Multi-Scale JPEG Compression Response Analysis for Image-Level Forgery Screening ​

Author: Sujith K Mandala
Published: 7/23/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2607.06615v2 Announce Type: replace-cross Abstract: Image forgery detection is a critical task in digital forensics, yet many deep-learning localization approaches are typically GPU-accelerated and computationally heavier than handcrafted screening methods. We propose a lightweight, interpreta...

📖 Read original article


261. Is Randomness Necessary for Adaptive Data Analysis? ​

Author: Edith Cohen, Haim Kaplan, Yishay Mansour, Shay Sapir, Uri Stemmer
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CR, cs.DS, cs.LG

arXiv:2607.07085v2 Announce Type: replace-cross Abstract: The Adaptive Data Analysis (ADA) problem formalizes the challenge of preventing false discovery and overfitting when a dataset is repeatedly reused. Formally, our input is a dataset containing $n$ i.i.d.\ samples from an unknown distribution ...

📖 Read original article


262. From Classification to Localization and Clinical Validation: Large-Scale Development of a Deep Learning System for Thoracic Disease Detection on Chest Radiographs in Thailand ​

Author: Isarun Chamveha, Tretap Promwiset, Napat Wanchaitanawong, Trongtum Tongdee, Pairash Saiviroonporn, Warasinee Chaisangmongkon
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.09305v2 Announce Type: replace-cross Abstract: Chest radiography (CXR) remains the most widely used thoracic imaging modality, yet expert interpretation is constrained by a severe shortage of radiologists in Thailand and across Southeast Asia. Local adaptation of deep learning models to T...

📖 Read original article


263. NexForge: Scaling Agent Capabilities through Requirement-Driven Task Synthesis for LLMs ​

Author: Jiarong Zhao, Zhikai Lei, Zhiheng Xi, Rui Zheng, Hang Yan, Jie Zhou, Qin Chen, Liang He
Published: 7/23/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2607.14186v4 Announce Type: replace-cross Abstract: Scaling executable agent training data for LLM post-training is bottlenecked by substrate-bound methods that tie task generation to predefined tools, repositories, or skill graphs: expanding coverage requires manual substrate engineering, eac...

📖 Read original article


264. Non-Asymptotic Variational Learning for Monotone Nonlinear Multiscale Elliptic Equations: Scale-Robust Primal-Dual Bounds and Strong-Form Statistical Ill-Conditioning ​

Author: Ronald Katende
Published: 7/23/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.AP

arXiv:2607.15702v2 Announce Type: replace-cross Abstract: We develop a non-asymptotic approximation, sampling, and finite-iteration optimization theory for variational physics-informed approximation of uniformly monotone nonlinear multiscale elliptic equations. For boundary-compatible neural feature...

📖 Read original article


265. Identity-Paired Progressive Depth Training: When Trainability Persists Beyond Expressibility ​

Author: Athanasios Hadjidimoulas, Tirthak Patel, Anastasios Kyrillidis
Published: 7/23/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, math.OC

arXiv:2607.16800v2 Announce Type: replace-cross Abstract: Variational Quantum Algorithms (VQAs) are a leading paradigm for near-term quantum computing, yet their training suffers from sensitivity to circuit depth, initialization, and landscape pathologies such as barren plateaus. We study \emph{prog...

📖 Read original article


266. Position: The Inevitable Transition to Machine Learning in Quantum Chemistry ​

Author: Karen Sargsyan, Chao-Ping Hsu
Published: 7/23/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, quant-ph

arXiv:2607.18281v2 Announce Type: replace-cross Abstract: Finding exact solutions to the quantum many-body problem is computationally intractable (QMA-hard). Traditional approximations for electrons in an atom or molecule -- density functional theory and wavefunction methods -- have been indispensab...

📖 Read original article


267. Enhanced Neural Quantum State via Annealed Gradient Descent ​

Author: Shiwei Zhou, Yiming Huang, Xiao Yuan, Xiaoxia Cai
Published: 7/23/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, physics.chem-ph, physics.comp-ph

arXiv:2607.18865v2 Announce Type: replace-cross Abstract: Neural quantum states offer expressive representations of quantum many-body wave functions, yet their practical accuracy can be limited by stochastic optimization rather than representational capacity. Here we identify a finite-sample instabi...

📖 Read original article


268. Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing ​

Author: Xinjie Zhang, Peng Zhang, Shicheng Zheng, Jinghao Guo, Zhaoyang Jia, Yifei Shen, Xun Guo, Yuxuan Luo, Jiahao Li, Wenxuan Xie, Fanyi Pu, Xiaoyi Zhang, Kaichen Zhang, Zongyu Guo, Tianci Bi, Dongnan Gui, Zhening Liu, Zimo Wen, Zihan Zheng, Senqiao Yang, Xiao Li, Jinglu Wang, Bin Li, Yan Lu
Published: 7/23/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MM, eess.IV

arXiv:2607.19064v2 Announce Type: replace-cross Abstract: Large-scale visual generators are increasingly capable but costly to train, fine-tune, and deploy. We introduce Mage-Flow, a compact 4B-scale generative stack for efficient text-to-image generation and instruction-based image editing. The sta...

📖 Read original article