Skip to content

arXiv cs.LG - 2026-08-17 ​

212 items collected.


1. L-FNO: Lorentzian Fourier Neural Operator for Stochastic Event Dynamics ​

Author: Songhee Kang, Jihoon Kang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.13562v1 Announce Type: new Abstract: Modern operational systems face uncertainty even in routine conditions, where rare, bursty, and self-exciting events emerge from both exogenous covariates and endogenous event dynamics. Standard neural operators are typically trained as regression-styl...

📖 Read original article


2. Don't Claim Benchmark-Oriented Optimization Improves General Coding Capability -- Diverse Evaluation Is Required ​

Author: Egor Shibaev, Vera Kudrevskaia, Timur Galimzyanov, Mikhail Evtikhiev, Ana Terna, Rastislav Rabatin, Timur Kudashev, Timofey Bryksin, Arina Puchkova, Patrik Bartak, Egor Bogomolov, Sergey Titov
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE

arXiv:2608.13566v1 Announce Type: new Abstract: Post-training papers, model cards, and blog posts often treat scores on a small set of coding benchmarks (e.g., SWE-bench and LiveCodeBench) as evidence of broad coding capability, both for research artifacts and user-facing systems. We argue that opti...

📖 Read original article


3. Robust XGBoosting for Regression ​

Author: Iris Arag'on Mladosich, Christophe Croux
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, stat.CO, stat.ML

arXiv:2608.13590v1 Announce Type: new Abstract: XGBoost is a very popular and powerful method for prediction. It iteratively fits simple decision trees to the residuals of the previous step. An efficient and scalable implementation is available. The standard loss function for XGBoost is the quadrati...

📖 Read original article


4. Training-Free Knowledge Transfer Across Model Scales through Activation-Guided Pruning ​

Author: Jiahe Fan, Si Chen, Yinghao Hou, Aiyuan Zhang, Hong Xie
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.13596v1 Announce Type: new Abstract: Heterogeneous model fusion seeks to combine models that differ in tasks, initializations, architectures, or scales. We study an underexplored cross-scale setting: improving a small recipient language model with a stronger donor despite substantial arch...

📖 Read original article


5. Hard Cases, Bad Labels: Testing Error Exposure and Error Location in Uncertainty Sampling Under Bounded Label Noise ​

Author: John Myron Uy
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13601v1 Announce Type: new Abstract: Active learning can reduce labeling cost by selecting informative examples, but the most uncertain examples may also be the hardest to label correctly. This study tests whether uncertainty sampling fails because it acquires more corrupted labels or bec...

📖 Read original article


Author: A. Quadir, A. Rahaman, Mushir Akhtar, M. Tanveer
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13628v1 Announce Type: new Abstract: Random vector functional link (RVFL) networks are lightweight and fast neural models that offer efficient training and strong generalization through randomized hidden-layer weights and direct input-output connections. However, conventional RVFL models ...

📖 Read original article


7. Contrastive Learning for Interpretable Anomaly Detection at Collider Experiments ​

Author: Haoyi Jia, Sagar Addepalli, Julia Gonski
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, hep-ex, hep-ph

arXiv:2608.13652v1 Announce Type: new Abstract: Generic event-level anomaly detection for collider physics has two recurring problems: anomaly scores are hard to interpret, and they correlate strongly with energy scale and object multiplicity. We present Organized Representation via Contrastive lear...

📖 Read original article


8. The Query Knows What to Forget: A Second Erase Direction for Linear Attention ​

Author: Dhruman Gupta, Aritra Das, Debayan Gupta
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13668v1 Announce Type: new Abstract: Linear attention keeps a state of fixed size. At long context, many stored items share this state, and interference between them degrades retrieval. Gated DeltaNet-2 (GDN-2), like every delta-rule model before it, derives its erase vector from the key ...

📖 Read original article


9. From BERT to Frontier Agents: Eight Years of Language-Model Progress, the Collapse of the Capability-Cost Curve, and the Rise of Task-Targeted Models ​

Author: Pranav Kumar Kaliaperumal
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.13675v1 Announce Type: new Abstract: Between October 2018 and July 2026 AI models progressed from simple systems like BERT to massive agents that solve complex math and write software. The ability to resolve real coding issues improved by nearly six times per year since late 2024. During ...

📖 Read original article


10. EEG-PRISM: Physiologically-Grounded Interpretability of Predictions by EEG Foundation Models ​

Author: Deeksha M Shama, Punnisa Amornsirikul, Archana Venkataraman
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13676v1 Announce Type: new Abstract: Objective: Foundation models represent the next advancement in AI for EEG analysis; however current explainable AI techniques provide attribution scores in the time-channel input space, which is mismatched to clinical intuition about EEG. Thus, there i...

📖 Read original article


11. SAGE: Surrogate-gradient Adaptation via Attention-Guided Entropy for Spiking Transformers ​

Author: Kiran Nair, Rodrigue Rizk, KC Santosh
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.NE

arXiv:2608.13702v1 Announce Type: new Abstract: Spiking neural networks (SNNs) offer an energy-efficient alternative to conventional deep neural networks by exploiting sparse event-driven computation, but their training remains challenging because the non-differentiable spike function requires surro...

📖 Read original article


12. Capacity-Dependent Effects of Data Selection for Reasoning ​

Author: Cuong Dang, Hoang Anh Just, Ruoxi Jia
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.13721v1 Announce Type: new Abstract: In reasoning supervised fine-tuning, candidate responses for the same instruction can differ substantially in how well they match the student's current distribution. Recent likelihood-based response selection methods suggest that responses closer to th...

📖 Read original article


13. The Integer Alibi: Localizing Cross-Kernel Divergence in INT8-Quantized LLM Inference ​

Author: Teng-Ruei Chen
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13756v1 Announce Type: new Abstract: Two GPU kernels implementing the same scaled INT8 GEMM interface are usually treated as interchangeable. We test that assumption: holding the checkpoint, prompts, hardware, inference engine, decoding, and quantization configuration fixed, we swap only ...

📖 Read original article


14. CutClean: Neural Network Pruning for Privacy-Preserving Inference ​

Author: Leonardo Magliolo, Vito Paolo Pastore, Giuseppe Valenzise, Enzo Tartaglione
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.13773v1 Announce Type: new Abstract: Neural networks are increasingly deployed in high-stakes applications with growing privacy leakage concerns. We show that this privacy leakage can occur even in the absence of representation imbalances that lead to traditional dataset biases. This pose...

📖 Read original article


15. PPAPlace: Differentiable Cross-Stage Objectives for Chip Placement Optimization ​

Author: Ruogu Chen, Jie Han
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR

arXiv:2608.13790v1 Announce Type: new Abstract: Macro placement significantly affects a chip's post-route performance, power, and area (PPA). Most placement methods optimize half-perimeter wirelength (HPWL) as the primary objective. However, recent benchmarking shows a near-zero correlation between ...

📖 Read original article


16. Recent Advances in Deep Learning-Based Drug-Target Binding Affinity Prediction ​

Author: Jafin Khan, Md Hossain Shuvo
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13797v1 Announce Type: new Abstract: Computational approaches to drug discovery involve multiple sub-problems, and among them, drug-target binding affinity prediction plays an important role. Despite recent advances, accurately predicting binding affinity remains an open research area. Th...

📖 Read original article


17. Dynamic Multi-Depot Vehicle Routing with Online Requests: Event-Driven Transformer--DRL and Rolling-Horizon Benchmarking ​

Author: Faezeh Ardali, Gerald M. Knapp
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13799v1 Announce Type: new Abstract: This paper presents an event-driven learning and benchmarking framework for the Dynamic Multi-Depot Vehicle Routing Problem with progressively revealed requests and evolving vehicle states. Masked MLP and Transformer policies are trained through behavi...

📖 Read original article


18. Stochastic Control Policies for Robust Molecular Transition Path Sampling ​

Author: Jingqian Liu, Yu-Hsiang Wang, Yanru Qu, Ge Liu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13800v1 Announce Type: new Abstract: Transition path sampling (TPS) aims to efficiently generate rare molecular transition trajectories between metastable states and is essential for understanding biomolecular mechanisms. Beyond traditional molecular dynamics (MD)-based sampling, machine ...

📖 Read original article


19. HI-MeshGraphNets: Efficient and Accurate Mesh-based Physics Learning with Hierarchical Multi-scale Graph Neural Networks ​

Author: SiHun Lee, Dong-Hyuk Park, Taesoo Bang, Seung-Hoon Kang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13827v1 Announce Type: new Abstract: Machine-learned physical surrogate models have become promising alternatives to mesh-based numerical solvers. Among them, graph neural networks (GNNs) are well suited for representing simulation meshes and learning nodal state evolution through message...

📖 Read original article


20. Federated Prompt Learning: A Unified Framework, Empirical Analysis, and Future Directions ​

Author: Qinglin Yang, Chen Qiu, Hongyuan Zhang, Pengdeng Li, Yuan Liu, Zhihong Tian
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC

arXiv:2608.13844v1 Announce Type: new Abstract: Large language models (LLMs) have become core components of cloud-based intelligent services in academia and industry, yet their training and deployment are hindered by high computational costs, data centralization, and privacy concerns. Federated lear...

📖 Read original article


21. Geometric Filtering of LLM-Generated Samples for Few-Shot Text Classification ​

Author: Benjam'in Schindler, Gonzalo A. Ruz
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.13866v1 Announce Type: new Abstract: Large language models (LLMs) can generate synthetic training data for text classification, but the quality of generated samples is heterogeneous: some fall in correct class regions of the embedding space while others land in peripheral or cross-class z...

📖 Read original article


22. Variation Brownian Kernel Ladders ​

Author: Mahdi Mohammadigohari
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13882v1 Announce Type: new Abstract: Claims about the benefit of depth depend on the complexity assigned to a representation. We introduce the \emph{Variation Brownian Kernel Ladder} (VBKL), a path-atomic function-space framework that separates nonlinear recursive dictionary construction ...

📖 Read original article


23. Fashion Outfit Generation via Unified Sequential Composition Models ​

Author: Kaicheng Pang, Xingxing Zou, Ruohan Xu, Waikeung Wong
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13888v1 Announce Type: new Abstract: The task of synthesizing stylistically coherent fashion outfits from massive item libraries, known as fashion outfit generation, remains a non-trivial challenge, primarily due to the non-monotonic and implicit nature of aesthetic compatibility, coupled...

📖 Read original article


24. MedMix: Specialization-Consistent Federated Sparse MoEs under Modality Heterogeneity ​

Author: Adiba Orzikulova, Dong Min Kim, Jaehong Yoon, Sung-Ju Lee
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13911v1 Announce Type: new Abstract: Federated multimodal medical AI faces modality heterogeneity at both the client and sample levels: clients may systematically lack access to specific modality types, while individual records within the same client may contain different partial modality...

📖 Read original article


25. Hybrid Quantum-inspired Kolmogorov-Arnold Networks for Privacy-Aware Federated Biosignal Learning ​

Author: Chun-Hua Lin, Samuel Yen-Chi Chen, Yu-Chao Hsu, Kuo-Chung Peng, Jiun-Cheng Jiang, Chi-Sheng Chen, Tai-Yue Li, Nan-Yow Chen, En-Jui Kuo, Hsi-Sheng Goan
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC, cs.ET, quant-ph

arXiv:2608.13914v1 Announce Type: new Abstract: Electrocardiogram (ECG) recordings are sensitive biomedical data, limiting the ability of hospitals and wearable devices to share raw signals for centralized model training. Federated learning addresses this practical privacy constraint by enabling col...

📖 Read original article


26. High-dimensional nonparametric changepoint detection via low-rank degree-two density projection ​

Author: Guoqing Zhang, Zhaixin Chen
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.13922v1 Announce Type: new Abstract: Detecting distributional changes in high dimension is difficult when neither the pre-change nor post-change density is parametrically specified. We introduce a representation-based approach that retains all degree-at-most-two density information while ...

📖 Read original article


27. CForce: Boosting Parallel Decoding for dLLMs via Consistency Forcing ​

Author: Yuji Ren, Chenkai Xu, Zhuocheng Gong, Jianguo Li, Zhijie Deng
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.13925v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) accelerate language generation by predicting multiple masks in a single forward pass. However, existing dLLMs can suffer from unreliable predictions in early denoising stages under aggressive parallelism strategi...

📖 Read original article


28. Post-training Quantization for Hybrid Iterative Generative Models ​

Author: Jing Gao, Junyi Wu, Wei Wang, Yan Yan, Yao Zhao
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13932v1 Announce Type: new Abstract: Iterative Generative Models (IGMs) span autoregressive and diffusion paradigms, and hybrid variants that couple them can achieve remarkable image-generation fidelity. However, their iterative inference incurs substantial computational overhead, making ...

📖 Read original article


29. Probabilistic indirect models for undrained shear strength: addressing significant data missing and variability with advanced imputation and machine learning techniques ​

Author: Haibin Xiong, Shaoheng Dai, Peng Lan, Xuzhen He, Chenxi Tong, Sheng Zhang, Daichao Sheng
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.DB

arXiv:2608.13934v1 Announce Type: new Abstract: Accurate prediction of undrained shear strength (su) is crucial for geotechnical design, but is often hampered by substantial uncertainty in traditional empirical methods. This study uses the CLAY/10/7490 global database to develop probabilistic indire...

📖 Read original article


30. Polar Code Based Federated Learning: Convergence Analysis and Resource Allocation ​

Author: Han Xiao, Wei Kang, Nan Liu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13961v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data; however, it faces significant communication bottlenecks and channel impairments in practice. Conventional network layer treatments either ...

📖 Read original article


31. QUASAR: Lowering the Loss Floor of Quantization-Aware Training with Loss-Aware Reconstruction ​

Author: Vincent Counathe, Ben Athiwaratkun, Christopher De Sa, Tianyi Zhang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, stat.ML

arXiv:2608.13966v1 Announce Type: new Abstract: As large language model inference shifts toward lower precision, post-training quantization (PTQ) becomes increasingly brittle, making quantization-aware training (QAT) essential for preserving model quality. However, QAT computes the loss and surrogat...

📖 Read original article


32. Identifiability and Order-Dimension Limits of In-Context Learning on Partial Orders ​

Author: Faizanuddin Ansari, Debanjan Dutta, Swagatam Das
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14004v1 Announce Type: new Abstract: In-context learning is commonly formalized as inference from examples of a function. Partial orders instead combine transitivity, antisymmetry, and incomparability, so a finite prompt may not determine a queried comparison. We develop a theory of in-co...

📖 Read original article


33. When Does More Correct Data Hurt? Insertion-Stability and the Limits of Dimension-Based Theory ​

Author: Joseph Sankoorikal Johny
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.14020v1 Announce Type: new Abstract: Adding data known to be correct ought to be safe. Not always. Larsen, Pabbaraju and Shetty model the failure with a monotone adversary, which reads an i.i.d. training sample and may append as many further examples as it likes, provided the target hypot...

📖 Read original article


34. Adversarial Learning of Classifier-Free Guidance Schedules ​

Author: Ashwini Pokle, Alexandre Galashov, Arnaud Doucet, Mauricio Delbracio, Valentin De Bortoli
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14038v1 Announce Type: new Abstract: Modern text-to-image diffusion models rely on classifier-free guidance (CFG) to achieve high image fidelity and text alignment. However, CFG typically applies a static, global scale across all timesteps, samples, and conditions -- a choice that is gene...

📖 Read original article


35. Model-agnostic Retrieval-Augmented Extended Forecasting for time series ​

Author: Juan Pablo Villa Serna, Rohan Asthana, Vasileios Belagiannis
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14054v1 Announce Type: new Abstract: Time series forecasting with pretrained foundation models has demonstrated strong zero-shot capabilities. However, achieving optimal performance on time series with short or negligible historical data in domain-specific applications typically requires ...

📖 Read original article


36. When Denoising Hurts: Rethinking the Terminal Step of Diffusion Time Series Forecasters -- Extended Version ​

Author: Dat Nguyen-Cong, Luong Tran, Tung Kieu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14067v1 Announce Type: new Abstract: Diffusion models offer a natural way to model uncertainty in time series forecasting, yet their iterative sampling process is often treated as a uniformly beneficial refinement procedure. Our study challenges this view by examining how forecast quality...

📖 Read original article


37. Resource-Adaptive Primal-Dual Learning for One-Warehouse Multi-Store Systems with Censored Demand ​

Author: Jiameng Lyu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.14096v1 Announce Type: new Abstract: The one-warehouse multi-store (OWMS) system is a fundamental inventory network in which a nonreplenishable warehouse allocates shared stock across multiple stores over time. Existing OWMS learning policies are built around a fixed target calibrated to ...

📖 Read original article


38. Sequence prediction under a lying oracle ​

Author: Puspabeethi Samanta, Nikhil Karamchandani, Jayakrishnan Nair
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT

arXiv:2608.14102v1 Announce Type: new Abstract: We consider the problem of sequential prediction of an $m$-ary sequence, where at each epoch, (i) the environment selects an outcome from an $m$-ary alphabet, (ii) the learner selects a probability distribution over the same alphabet (unaware of the ou...

📖 Read original article


39. Forecast Collapse in Time-Series Foundation Models ​

Author: Shu Wan, Miles Ma, Hank Zhu, Guangqi Liu, Stephen Wang, Qingsong Wen, Huan Liu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE, stat.AP, stat.ML

arXiv:2608.14106v1 Announce Type: new Abstract: When forecasting hourly returns for 1,000 US equities, we observe an unexpected phenomenon: predictions become nearly flat and show poor stock ranking, as measured by cross-sectional correlation. We call this forecast collapse. Surprisingly, the phenom...

📖 Read original article


40. Learning to Run Power Networks: Effective AlphaZero-inspired Topological Control ​

Author: Lukas Zetto, Benjamin Sch"afer, Qiong Huang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14114v1 Announce Type: new Abstract: As the integration of volatile renewable energy sources increases the strain on modern power grids, the use of Reinforcement Learning (RL) for autonomous topological reconfiguration has emerged as a promising research field to keep strained grids stabl...

📖 Read original article


41. From Fixed Grids to Moving Particles:A Transferable Latent Operator for Fluid Dynamics ​

Author: Meng Li, Chuqi Chen, Zhengqing Gao, Xi Zhou, Xiao Sun, Yang Xiang, Huaxi Huang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GR

arXiv:2608.14120v1 Announce Type: new Abstract: Lagrangian modeling is vital to fluid dynamics, as it characterizes particle transport and complements the Eulerian description.However, Lagrangian trajectories are less commonly available than Eulerian fields, while most neural operators are trained a...

📖 Read original article


42. Overcoming Shortcut Learning in Graph Neural Networks through Active Explanation Guidance ​

Author: Taraneh Younesian, Steve Azzolin, Antonio Longa, Francesco Ferrini, Vincenzo Marco De Luca, Stefano Teso
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14121v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) can solve prediction tasks by unintentionally exploiting shortcuts---that is, edges, nodes, and features that correlate with but are not causal for the prediction---which compromise their reliability in out-of-distribution ...

📖 Read original article


43. Smart routes: a system for development and comparison of algorithms for solving vehicle routing problems with realistic constraints ​

Author: Andrew Soroka, German Mikhelson, Alexander Mescheryakov, Sergey Gerasimov
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14140v1 Announce Type: new Abstract: The problem of route optimization with realistic constraints is becoming extremely relevant in the face of global urban population growth. While we are aware of approaches that theoretically provide an exact optimal solution, their application becomes ...

📖 Read original article


44. Deep Reinforcement Learning solution for pickup and delivery routing problems with time window and capacity constraints ​

Author: Andrew Soroka, Alex Meshcheryakov, Sergey Gerasimov
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14156v1 Announce Type: new Abstract: The task of constructing vehicles optimal routes for pickup and delivery of goods is one of most promising tasks in the context of global urban population growth. Although this kind of problems with small size can be solved by various classical approac...

📖 Read original article


45. Structure-Guided Spatiotemporal Attention Graph Neural Network for Traffic Flow Prediction ​

Author: Xuanmian He, Can Li, Wanjing Ma
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14177v1 Announce Type: new Abstract: Deep spatiotemporal models integrating graph convolutions and attention mechanisms have demonstrated excellent performance in network-level traffic flow prediction, owing to their exceptional ability to capture complex spatiotemporal dependencies. Desp...

📖 Read original article


46. Revisiting Energy-based Tabular Anomaly Detection: Energy and Reconstruction are Complementary ​

Author: Junichiro Niimi
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.14186v1 Announce Type: new Abstract: Tabular anomaly detection is dominated by classical density-proxy methods (Isolation Forest, OCSVM, LOF), reconstruction-based detectors (Autoencoders, VAEs), and modern non-parametric scorers (COPOD, ECOD, Deep SVDD), all of which approximate the inli...

📖 Read original article


47. KV Cache Compression Through the Lens of Transform Coding ​

Author: Hannah Laus, Claudio Mayrink Verdun, Hao Wang, Flavio du Pin Calmon, Felix Krahmer
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, eess.SP

arXiv:2608.14191v1 Announce Type: new Abstract: The key-value (KV) cache stores information from past tokens and is a major memory bottleneck in long-context inference. Existing quantization methods address this bottleneck by representing the KV cache uniformly with lower-precision data types and de...

📖 Read original article


48. MINT: A Universal Zero-Shot Predictor for Transaction Data ​

Author: Parameswaran Kamalaruban, Viktor Drobnyi, Maeve Madigan, Julia Rozanova, David Sutton, Stuart Burrell
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.14198v1 Announce Type: new Abstract: Banks analyse sequential financial transaction data to perform many tasks, including fraud prevention, credit risk assessment and offer personalization. To improve the predictive accuracy of these tasks, Payments Foundation Models encode transaction se...

📖 Read original article


49. Adaptive Protection for Evolutionary Feature Construction in Symbolic Regression with Application to Credit Classification ​

Author: Hengzhe Zhang, Qi Chen, Bing Xue, Lean Yu, Wolfgang Banzhaf, Mengjie Zhang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2608.14209v1 Announce Type: new Abstract: Evolutionary feature construction has shown strong promise in symbolic regression by automatically discovering informative transformations of input features that enhance a simple base learner. However, existing approaches often lack explicit mechanisms...

📖 Read original article


50. Training Fair Tabular Foundation Models ​

Author: Patrik Kenfack, Jesse C. Cresswell, Anthony L. Caterini, Samira Ebrahimi Kahou, Ulrich A"ivodji
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14211v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) have emerged as leading methods for tabular predictive tasks, leveraging in-context learning to predict on new data without task-specific training. Despite the increased use of TFMs in high-stakes decision-making, their...

📖 Read original article


51. Connected Subspace Clustering: Hardness, a Scalable Heuristic, and an Application to Sea Level Geodesy ​

Author: Johanna Hillebrand, Jan H"ockendorff, J"urgen Kusche, Kelin Luo, Heiko R"oglin, Melanie Schmidt, Christian Sohler, Bernd Uebbing
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14215v1 Announce Type: new Abstract: Constrained optimization extends classical optimization by integrating side information, making it widely applicable across scientific and engineering domains. Consider a setting where we measure variables at different physical locations. When grouping...

📖 Read original article


52. AutoSchema: Live Schema Grounding for Agentic Text-to-Sparql over Heterogeneous Knowledge Graphs ​

Author: Yiming Zhang, Koji Tsuda
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14228v1 Announce Type: new Abstract: Life science knowledge graphs make large collections of structured data available through SPARQL, but each resource uses its own schema, identifiers, and links. TogoMCP helps language model agents query these resources by providing curated Metadata Int...

📖 Read original article


53. Multi-Objective Bayesian Optimization for Model Merging ​

Author: Utkarsh Agarwal, Vamshi Bonagiri, Raul Astudillo, Monojit Choudhury
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14264v1 Announce Type: new Abstract: Model merging combines trained models directly in weight space, offering a compute-efficient alternative to additional fine-tuning. Selecting merge parameters is nevertheless difficult because downstream evaluations are expensive, gradients are unavail...

📖 Read original article


54. Convex losses and their applications to SVM, SVR, and Shallow Neural Networks ​

Author: Filippo Portera
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14288v1 Announce Type: new Abstract: We propose multiple new convex losses for SVM and Neural Networks, applied to binary classification tasks. While there are practical limitations in exploiting them with the dual SVM models, we are able to use them with SVM primal formulation and Neural...

📖 Read original article


55. Detecting Contaminated Code-Generation Prompt Batches via Influence Functions ​

Author: Francesco Quinzan, Noor Munir, Yishun Lu, Stephen Roberts
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14303v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for code generation, yet they remain vulnerable to prompts that elicit insecure implementations. Existing defenses typically rely on predefined threat models or known vulnerability patterns, limiting t...

📖 Read original article


56. Quantum Multi-Armed Bandits and Linear Bandits: Lower Bounds and Algorithms ​

Author: Maoli Liu, Zhuohua Li, John C. S. Lui
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, quant-ph

arXiv:2608.14319v1 Announce Type: new Abstract: We study quantum multi-armed bandits (QMAB) and quantum linear bandits (QLB) in the model of Wan et al. [2023], where the learner queries each arm or action through a quantum reward oracle or its inverse. Prior work gives algorithms over horizon $T$ wi...

📖 Read original article


57. Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling ​

Author: Michael Fore, Akshay Jain, Justin Downes, Rohan Pradhan, Duncan Botti
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14349v1 Announce Type: new Abstract: We present a training-free method for multi-modal trajectory prediction that achieves comparable accuracy to a 57M-parameter transformer while requiring no GPU and zero learned parameters. The method builds a transition table of historical state-to-nex...

📖 Read original article


58. Mind the Long Tail: Understanding the Difficulty of Delay Detection in Business Processes ​

Author: Keyvan Amiri Elyasi, Lukas Kirchdorfer, Heiner Stuckenschmidt
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14367v1 Announce Type: new Abstract: The early detection of delayed cases in business processes is a critical capability for organizations. Predictive process monitoring (PPM) supports this task by using historical event logs to predict the remaining time of ongoing cases, enabling timely...

📖 Read original article


59. Catching the Imposter: Self-Supervised Learning of Physical Coherence with Cross-Entity Feature Permutations ​

Author: Aleksei Rozanov, Arvind Renganathan, Vipin Kumar
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14372v1 Announce Type: new Abstract: Scientific data often describe entities whose features are jointly governed by the laws of physics, yet existing self-supervised learning (SSL) objectives largely ignore this physical coherence. We introduce imposter, a discriminative pretext task that...

📖 Read original article


60. Boosting Data Augmentation with Stochastic Weight Averaging ​

Author: Longde Huang, Axel Flinth, Jan E. Gerken
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14373v1 Announce Type: new Abstract: The symmetries of a learning task have become an important factor in designing modern deep learning solutions. Data augmentation is a straightforward and effective way of incorporating symmetries into a generic neural network. Recent results show that ...

📖 Read original article


61. DeaMoE: Efficient MoE Structure for Fast Small-Batch Decoding ​

Author: Zewen Jin, Shen Fu, Zeping Duan, Shannon Wang, Weihao Wu, Chengjie Tang, Congkun Ai, Ping Gong, Zijian Dai, Youhui Bai, Cheng Li
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14385v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have been widely adopted in real-time interactive applications such as coding assistants, real-time audio-video interaction systems. To meet the extremely low response latency requirements of these scenarios, practitione...

📖 Read original article


62. CytoBERT: A Foundation Model for Cytometry Data ​

Author: Syed Abdul Haseeb Qadri, Bjarne C. Hiller, Felix Blanke, Vanja Sophie Cangalovic, Kutalm{\i}\c{s} Co\c{s}kun, Amin Mirzaei, Tom Siegl, Sebastian Bader, Thomas Kirste, Martin Becker
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14414v1 Announce Type: new Abstract: Cytometry measures the complex characteristics of single cells (e.g., counts and protein expression of immune cells) and is widely used across immunological research and clinical settings. However, cytometry data is highly heterogeneous and unstandardi...

📖 Read original article


63. More Correct Mass, Worse Answers: Why Power Sampling Can Fail and How to Fix It ​

Author: Haohui Yang, Jiaxing Sun, Xiujun Ma
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14420v1 Announce Type: new Abstract: Power Sampling sharpens a language model's distribution over complete generation trajectories, offering a verifier-free way to improve reasoning at inference time. It also has the potential to serve as a general-purpose front end for a broad range of d...

📖 Read original article


64. Designing Reinforcement Learning for Diffusion Models: A Unified Path-Space View ​

Author: Yixian Xu, Yuanrui Zhang, Shengjie Luo, Liwei Wang, Di He
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, stat.ML

arXiv:2608.14430v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training provides a direct way to align diffusion models with human preferences and task-specific rewards. However, current RL algorithms for diffusion models remain fragmented: reverse-trajectory methods rely on discre...

📖 Read original article


65. Designing Compact Neural Architectures via Neuron Gating and Mixed Activation ​

Author: Abhishek Shukla, Ankur Sinha, Faiz Hamid
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14443v1 Announce Type: new Abstract: Neural Architecture Search (NAS) is naturally formulated as a bilevel optimization problem, where the upper-level optimizes the architecture using validation performance and the lower-level trains network parameters using training loss. However, NAS is...

📖 Read original article


Author: Abhishek Shukla, Ankur Sinha, Faiz Hamid
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14472v1 Announce Type: new Abstract: Neural Architecture Search (NAS) aims to automate neural network architecture design, reducing reliance on human expertise. Among the various NAS methods, differentiable NAS has gained prominence due to its efficiency and accuracy compared to conventio...

📖 Read original article


67. Approximate Muon with low-rank adapters ​

Author: Ben Anson, Conor Houghton, Edward Milsom
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14492v1 Announce Type: new Abstract: The Muon optimizer shows clear benefits versus alternatives when pretraining neural networks. However, it is used less frequently for parameter-efficient fine-tuning (PEFT). One potential reason is that the most common PEFT method, LoRA, does not natur...

📖 Read original article


68. Generating Benchmark Health Data Using a Tabular Diffusion Transformer ​

Author: Hao Yan, Lisa Pilgram, Dan Liu, Linglong Kong, Fida Dankar, Khaled El Emam
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14496v1 Announce Type: new Abstract: Cross-Tabular Data Generation (CTDG) seeks to learn a generative model from multiple heterogeneous tables and produce new synthetic tabular datasets. However, existing synthetic tabular data generation methods are largely restricted to single-input-tab...

📖 Read original article


69. Rollplex: Cross-Phase GPU Spatial Sharing for Vision Language Model Post-Training ​

Author: Hanfeng Lu, Tianyu Feng, Suyi Li, Yuheng Zhao, Wei Gao, Shaopan Xiong, Ju Huang, Siran Yang, Jiamang Wang, Lin Qu, Wei Wang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2608.14498v1 Announce Type: new Abstract: Vision-language models (VLMs) enable embodied agents to reason and act from visual observations and language instructions. Reinforcement learning (RL) post-training enhances these capabilities using task feedback, but current on-policy RL runtimes exec...

📖 Read original article


70. RecipeNet: A Hierarchical Transformer for Recipe Data ​

Author: Pin-Yen Huang, Sachin Chhabra, Prasanth Sai Gouripeddi, Abhinav Kumar, Baoxin Li
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14505v1 Announce Type: new Abstract: Recipe data arises in domains such as materials synthesis, pharmaceutical formulation, and industrial manufacturing, where procedures are represented as ordered sequences of steps containing heterogeneous structured fields. Existing tabular learning me...

📖 Read original article


71. Modular Cognitive Architecture Emerges in Large Language Models ​

Author: Pengrui Han, Jacob Andreas, Evelina Fedorenko, Andrea Gregor de Varda
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.13567v1 Announce Type: cross Abstract: The human brain exhibits a striking degree of functional specialization, with distinct networks supporting language, formal reasoning, reasoning about other minds, and reasoning about the physical world. Is this modular organization a fundamental pri...

📖 Read original article


72. Think in Latent, Explain in Language: Self-Explainable Latent Reasoning ​

Author: Dayuan Zhao, Shengcao Cao, Yu-Xiong Wang, Liang-Yan Gui
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.13570v1 Announce Type: cross Abstract: Latent reasoning has emerged as a powerful alternative to text-based Chain-of-Thought (CoT), offering significant gains in computational efficiency by compressing verbose reasoning into compact embeddings. However, compressing reasoning into the late...

📖 Read original article


73. Interactive Analysis of Global Explanations using Aggregated Class Activation Maps for Network Data ​

Author: Igor Cherepanov, David Sessler, Alex Ulmer, Felix Wagner, Throsten May, J"orn Kohlhammer
Published: 8/17/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.LG, cs.NI

arXiv:2608.13575v1 Announce Type: cross Abstract: Recent machine learning (ML) advances have demonstrated that deep learning (DL) achieves impressive results in different application domains, including the classification of computer network traffic to corresponding applications. However, the data fr...

📖 Read original article


74. BCIJelly: An integrated ecosystem for brain-computer interface research ​

Author: Liyuan Han, Xinrui Yang, Tianyu Zheng, Qizhi Yang, Yitao Qin, Liang Chen, Qinglai Wei, Binjie Hong, Xinhe Zhang, Rui Xiong, Yong Gu, Mu-ming Poo, Bo Xu, Chengyu Li, Tielin Zhang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.HC, cs.LG, q-bio.NC

arXiv:2608.13576v1 Announce Type: cross Abstract: Brain-computer interface (BCI) research relies on multistage computational pipelines, yet progress remains constrained by fragmented data formats, heterogeneous decoder implementations and hardware-specific deployment toolchains, and researchers lack...

📖 Read original article


75. AI Evaluation Should Work With Humans ​

Author: Jan Kulveit, Gavin Leech, Tom'a\v{s} Gaven\v{c}iak, Raymond Douglas
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.13577v1 Announce Type: cross Abstract: This position paper argues that the dominant paradigm of AI evaluation (which focuses on superhuman autonomous performance and so implicitly targets the goal of replacing humans) is guiding AI development in the wrong direction. Instead, the AI commu...

📖 Read original article


76. BCMT: Blockwise Causal Memory Transformer ​

Author: Rachid Arezki
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.13578v1 Announce Type: cross Abstract: Transformer architectures rely on dense self-attention to model long-range dependencies, but this mechanism exhibits quadratic complexity with respect to sequence length. We introduce BCMT (Blockwise Causal Memory Transformer), an architecture for lo...

📖 Read original article


77. From Prediction to Intervention: Personalized Meal-Level Glucose Regulation via an LLM Agent ​

Author: Mingyu Huang, Weiqing Min, Ying Jin, Yilin Wang, Shuqiang Jiang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.LG

arXiv:2608.13581v1 Announce Type: cross Abstract: Personalized glucose regulation remains a central yet unresolved challenge in precision nutrition, as postprandial glucose response varies substantially across individuals. Existing approaches based on glycemic indices fail to adequately account for ...

📖 Read original article


78. Continual Evolution Strategies in Control Tasks ​

Author: Nicola Pitzalis, Eleni Nisioti, Antonio Carta, Davide Bacciu, Andrea Cossu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2608.13600v1 Announce Type: cross Abstract: We study Evolution Strategies (ES) for continual control, where agents must adapt to changing tasks without forgetting previous ones. On sequential MuJoCo locomotion tasks, naive ES suffers from severe catastrophic forgetting. Replay substantially im...

📖 Read original article


79. MobileMem: Learning from a Year of Mobile Experiences ​

Author: Xinle Deng, Yida Xue, Xiangyuan Ru, Haoming Xu, Shuofei Qiao, Mengru Wang, Yijun Chen, Buqiang Xu, Chen Jiang, Yuchen Eleanor Jiang, Lizhong Wang, Jianfeng Wang, Li Zeng, Haofen Wang, Guilin Qi, Huajun Chen, Ningyu Zhang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA, cs.MM

arXiv:2608.13606v1 Announce Type: cross Abstract: The next generation of AI agents is increasingly moving beyond systems that answer isolated questions toward persistent personal assistants that can understand, remember, and continuously learn from users' experiences. Such assistants require long-te...

📖 Read original article


80. No Universal Signal Predicts Sample-Level LLM Regression under Version Updates ​

Author: Jia Sheng, Yiwei Lu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.13607v1 Announce Type: cross Abstract: Frontier LLMs are updated frequently and typically outperform their predecessors in aggregate. But aggregate gains say little about individual samples: an update can still cause sample-level regression, where a response correct under the old model be...

📖 Read original article


81. Evaluating Agentic Learning Harness Capabilities Without Labels via the Scaling Hypothesis ​

Author: Aryan Luthra, Kshitij Jain, Siddharth Arya, Bobby Filar, Anna Bertiger
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG

arXiv:2608.13608v1 Announce Type: cross Abstract: Agentic "Continual Learning Harnesses", systems that pair an LLM with retrieval or memory to improve from feedback without retraining, have shown growing value in cybersecurity. But their value is conventionally measured by gains against labeled benc...

📖 Read original article


82. VoiceDesigner: Text-to-Voice Generation and Editing via Unified Diffusion Modeling and Data Augmentation ​

Author: Jiarui Hai, Karan Thakkar, Ke Chen, Yunyun Wang, Jiaqi Su, Rithesh Kumar, Mounya Elhilali, Zeyu Jin
Published: 8/17/2026, 4:00:00 AM
Categories: eess.AS, cs.LG

arXiv:2608.13613v1 Announce Type: cross Abstract: Recent breakthroughs in generative models have made text-to-voice generation (TTV) possible, enabling the synthesis of speech directly from textual voice descriptions. However, existing systems face two key challenges. First, they struggle to generat...

📖 Read original article


83. Adjacency-Based Spectral Proxy Control of Mobile Communication Agents ​

Author: Mariana del Castillo, Federico Larroca
Published: 8/17/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.MA, cs.SY, eess.SY

arXiv:2608.13616v1 Announce Type: cross Abstract: We consider a heterogeneous mobile-agent network composed of uncontrolled task agents and controllable communication agents. The objective is to reposition communication agents online as task agents move. Since throughput-based objectives are general...

📖 Read original article


84. Reward Machines for Signal Temporal Logic ​

Author: Alper Kamil Bozkurt, Shangtong Zhang, Yuichi Motai
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.RO

arXiv:2608.13625v1 Announce Type: cross Abstract: Signal temporal logic (STL) provides a formal language for specifying real-time properties of real-valued observations, along with a quantitative robustness score for monitoring satisfaction. Control synthesis from STL specifications is of interest s...

📖 Read original article


85. A Calibrated Test of Internal Action Maps: State Signals Without Global Affine Closure ​

Author: Dekun Yang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.13626v1 Announce Type: cross Abstract: A hidden state signal can be decodable or causally usable without supporting a reusable action map. We test whether action maps fitted without a source reach its natural post-action activation and compose. We organize the tests as an evidence lattice...

📖 Read original article


86. Learning Unsteady Aneurysm Hemodynamics with Physics-Informed DeepONets ​

Author: Oscar L. Cruz-Gonzalez, Val'erie Deplano, Badih Ghattas
Published: 8/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, physics.flu-dyn

arXiv:2608.13629v1 Announce Type: cross Abstract: Clinically actionable, patient-specific hemodynamic assessment, specifically wall shear stress, vortex structure and pressure distributions, is critical for determining risky or unfavorable evolution in Abdominal Aortic Aneurysms (AAA). While Physics...

📖 Read original article


87. Unknown Unknowns: Model Misspecification in Machine Learning for Physics ​

Author: Juan Cruz-Martinez, Carolina Cuesta-Lazaro, Alexander Held, Michael Kagan
Published: 8/17/2026, 4:00:00 AM
Categories: physics.data-an, astro-ph.CO, astro-ph.GA, cs.LG, hep-ex, hep-ph

arXiv:2608.13633v1 Announce Type: cross Abstract: Machine learning is now a central tool for solving inverse problems in particle physics and astronomy. Models are trained on simulation and deployed on real data, raising the question not just of whether they fit, but of whether they are wrong in way...

📖 Read original article


88. Consistent Model Chasing Is Minimax Optimal: The Exact Value of Scalar Adversarial Adaptive Control under Large Parametric Uncertainty ​

Author: Dimitar Ho
Published: 8/17/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC

arXiv:2608.13651v1 Announce Type: cross Abstract: We solve exactly a fundamental problem of adaptive control against adversarial disturbances: regulate the scalar system $x_{t+1} = ax_t + u_t + w_t$, $x_0=0$, $|w|_\infty \le 1$, where the constant pole $a \in [-\Delta, \Delta]$ is unknown in sign ...

📖 Read original article


89. What to Preserve, Where to Adapt: A Depth-Wise Analysis of Forgetting in Continual Gynecological Image Segmentation ​

Author: Amal Saqib, Tausifa Jan Saleem, Numan Saeed, Mohammad Yaqub
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.13660v1 Announce Type: cross Abstract: Medical image segmentation models are typically trained under the assumption that all data are available simultaneously. However, in clinical practice, datasets often arrive sequentially, requiring models to adapt continuously to evolving data distri...

📖 Read original article


90. hint$^2$: Hierarchical World Models for Inference-Time Temporal Logic Guidance ​

Author: Moritz Zoellner, Anastasios Manganaris, Ahmed H. Qureshi, Rohan Paleja
Published: 8/17/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.13678v1 Announce Type: cross Abstract: A central goal of robot learning is to enable robots to execute rich instructions specified at runtime. Large-scale language-conditioned policies have made substantial progress toward this goal, yet still struggle with temporal structure and safety c...

📖 Read original article


91. Language-Specific Gaps in AI Safety Training Datasets ​

Author: Chialuka Prisca-Mary Onuoha, Bright Etornam Sunu, Rashidat Sikiru
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CY, cs.LG

arXiv:2608.13695v1 Announce Type: cross Abstract: Large language model providers routinely cite multilingual safety benchmarks spanning a dozen or more languages as evidence that their models are safe for non-English-speaking users. We show that these collection-level coverage claims frequently do n...

📖 Read original article


92. GRPO Beyond English: A Large-Scale Study of GRPO in Non-English and Multilingual Settings ​

Author: Konstantin Dobler, Federico Scozzafava, Jonathan Janke, Mohamed Ali, Simon Lehnerer
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.13698v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR), often optimized with Group Relative Policy Optimization (GRPO), has become a central recipe for improving the reasoning capabilities of pretrained language models but current studies remain heavi...

📖 Read original article


93. TRUE-Colon: Exposing a Consistent Transfer Asymmetry in Real-Time Polyp Detection ​

Author: Sebastian Doerrich, Andreas Franz Schwab, Francesco Di Salvo, Shyam Nandan Rai, Hanh Huyen My Nguyen, Christian Ledig
Published: 8/17/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2608.13711v1 Announce Type: cross Abstract: Computer-aided detection (CADe) systems for colonoscopy promise to reduce clinical miss rates, yet reliable real-world deployment remains elusive. This translational gap stems in part from a structural flaw in model development: the reliance on curat...

📖 Read original article


94. Architecture and Affordances of PLAUD: Performative Latents and Unsupervised DDSP ​

Author: B{\l}a.zej Kotowski, Frederic Font
Published: 8/17/2026, 4:00:00 AM
Categories: cs.SD, cs.HC, cs.LG

arXiv:2608.13724v1 Announce Type: cross Abstract: PLAUD (Performative Latents and Unsupervised DDSP) is a neural synthesizer and Max for Live instrument for live electronic music, built on NoiseBandNet and trained on small personal sound corpora. We present its architecture, combining a variational ...

📖 Read original article


95. Building AI-Intensive Software with AI: Early Results and a Cautionary Tale on Measuring Development Cost ​

Author: Victor Barros de Miranda Neves, Kiev Santos da Gama, Vinicius Cardoso Garcia
Published: 8/17/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2608.13730v1 Announce Type: cross Abstract: Empirical reports on the true cost of AI-intensive software development remain scarce, and the few that exist are easy to get wrong in ways that never surface in the final number. We report early results from an ongoing case study: a six-person stude...

📖 Read original article


96. GALA: Generation-Aware Cross-Modal Alignment for Text-to-Time-Series Synthesis ​

Author: Haochen Zhang, Gengwei Zhang, Laura Yao, Nicholas Knoz, Tianlong Chen
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.13741v1 Announce Type: cross Abstract: Synthesizing time series from natural language is emerging as the most expressive form of controllable time series generation. However, existing text-conditioned generators either take caption embeddings frozen from off-the-shelf text encoders, or ad...

📖 Read original article


97. Does ISO-Grounded NFR Specification Improve LLM Code Generation? A Comparison of Rich and Structured Interventions against a Natural-Language Baseline ​

Author: Jo`ao Pedro Monteiro Pereira, Vinicius Cardoso Garcia
Published: 8/17/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2608.13742v1 Announce Type: cross Abstract: In LLM-based code generation, Non-Functional Requirements (NFRs) are often specified as terse one-line phrases. We ask whether grounding those specifications in ISO/IEC 25010 Quality Model, either as rich natural-language prose (NL-rich) or as struct...

📖 Read original article


98. Data-driven techniques for translational neuroscience and personalized neuro-health ​

Author: Vishal Subedi, Shashipraba N. K. Rajakaruna, Pratyusha Sarkar, Subhankar Chattoraj, Anjali Khasa, Siddhartha Nandy, Hamza Farooq, Animikh Biswas, Sanjay Chaudhuri, Asim K. Dey, Karuna Joshi, Christophe Lenglet, Ansu Chatterjee
Published: 8/17/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.LG, stat.AP, stat.ML

arXiv:2608.13749v1 Announce Type: cross Abstract: Neurodegenexrative diseases such as Alzheimer's disease and Parkinson's disease are diagnosed most reliably only after substantial, often irreversible, neuronal loss has already occurred, creating an urgent need for quantitative tools that can detect...

📖 Read original article


99. Amplified Does Not Mean Predictive: Reasoning Behaviors in Thinking Models ​

Author: Jean de Dieu Nyandwi, Leena Mathur, Yonatan Bisk, Robert Hawkins, Graham Neubig
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CV, cs.LG

arXiv:2608.13760v1 Announce Type: cross Abstract: Which reasoning behaviors are associated with correct answers in reasoning models, and does reasoning-oriented training amplify those behaviors? This distinction is important because reasoning-oriented training can make traces look more deliberative ...

📖 Read original article


100. From Passive Delegates to Strategic Negotiators: Reinforcing Social Reasoning in Small Language Models with SocialRL ​

Author: Wenyue Hua, Zachary Huang, Tyler Payne, Safoora Yousefi, Saleema Amershi, Asli Celikyilmaz
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2608.13787v1 Announce Type: cross Abstract: AI agents increasingly act on their users' behalf, handling tasks such as scheduling meetings, comparing offers, and haggling over prices. These principal-driven tasks routinely place the agent across from a counterpart (another user's agent, a selle...

📖 Read original article


101. On the Brittleness of Maximum Likelihood Estimation for Gaussian Process Hyperparameter Optimization ​

Author: Tyler R. Johnson, Kian Ben-Jacob, Christopher P. Muller, Ramin Bostanabad
Published: 8/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.13793v1 Announce Type: cross Abstract: Machine learning (ML) has become an indispensable part of modern engineering design workflows. A crucial step in training an ML model is the selection of the loss function which can be systematically formulated via various techniques such as maximum ...

📖 Read original article


102. What preferences can - and cannot - predict in multi-agent online learning ​

Author: Omar Abbadi, Rida Laraki, Panayotis Mertikopoulos
Published: 8/17/2026, 4:00:00 AM
Categories: cs.GT, cs.LG

arXiv:2608.13810v1 Announce Type: cross Abstract: We examine the interplay between ordinal, preference-based solution concepts in games and the long-run behavior of game dynamics, asking in particular to what extent the combinatorial data of a game -- its preference graph -- determine the outcomes o...

📖 Read original article


103. Trajectory Dynamics in Self-Supervised Learning Latent Space for Audio Deepfake Detection ​

Author: Tom'as Andrade Weber
Published: 8/17/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, cs.SD

arXiv:2608.13817v1 Announce Type: cross Abstract: Human speech production is constrained by physiology, giving rise to characteristic temporal structure on acoustic signals. We hypothesise that these constraints manifest as structured trajectory dynamics in the latent space of Self-Supervised Learni...

📖 Read original article


104. SPEAR: Structure Property Explainability with Attention Regularization ​

Author: Aditya Raghavan, Utkarsh Pratiush, Dalton A. Pearl, Jade Holliman Jr, Katharine Page, Philip D Rack, Sergei V Kalinin
Published: 8/17/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG

arXiv:2608.13826v1 Announce Type: cross Abstract: Machine learning is increasingly used to learn structure property relationships from spectroscopic and diffraction data, yet its adoption in materials discovery is often limited by poor interpretability of model predictions. Although attention mechan...

📖 Read original article


105. Engineering Signals of Human-AI Collaboration in the Agentic Coding Era: A Longitudinal Analysis of 33,228 Pull Requests from vLLM and SGLang with Implications for Biomedical AI Agents and Bioinformatics Pipeline Developmen ​

Author: Jiada Li, Xuesong Ye, Olamide Olowoniyi
Published: 8/17/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.ET, cs.HC, cs.LG

arXiv:2608.13884v1 Announce Type: cross Abstract: The rapid adoption of AI coding assistants and autonomous agentic development systems has coincided with major changes in the pace and structure of open-source software engineering. Yet empirical longitudinal evidence of these changes at the team lev...

📖 Read original article


106. Agentic Transaction: Towards ACID-Compliant Agent Systems ​

Author: Zhaoyan Sun, Xiaoxiao Wang, Guoliang Li
Published: 8/17/2026, 4:00:00 AM
Categories: cs.DB, cs.AI, cs.CL, cs.LG

arXiv:2608.13900v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from conversational assistants into autonomous systems that execute long-horizon tasks through reasoning, tool use, code generation, and workspace manipulation. As agents increasingly operate over persis...

📖 Read original article


107. Deep Vision in Smart Manufacturing: MODERN Framework for Intelligent Quality Monitoring and Diagnosis ​

Author: Yicheng Kang, Yuling Jiao, Xin Geng, Mahesh Nagarajan
Published: 8/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.13937v1 Announce Type: cross Abstract: Smart manufacturing processes are often installed with a large number of sensors, imaging devices and computers, which not only enable instant communication across various modules of a production system but also aid in intelligent manufacturing manag...

📖 Read original article


108. Nanbeige4.2-3B on Apple Silicon: Fixing Deployment Bugs and Decreasing Looped Transformer Memory Overhead ​

Author: John T. Halloran
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.13987v1 Announce Type: cross Abstract: Nanbeige4.2-3B is a 3B-parameter agentic model built around a Looped Transformer (LT) that reuses one stack of layers for a second forward pass, adding effective depth without additional parameters. Evaluated on Apple Silicon (MPS), we identify five ...

📖 Read original article


109. Buy the Rumor, Sell the News: When Is News Priced In? ​

Author: Alireza Kargarzadeh, Nariman Khaledian, Navid Parvini, Sid Ghatak, Arman Khaledian
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-fin.ST

arXiv:2608.14014v1 Announce Type: cross Abstract: Two old market sayings hold that news is already priced in by the time it is published, and that the rumor is bought while the news is sold. Both place the price move associated with a piece of news before and at publication rather than after it. Whe...

📖 Read original article


110. Emergent Models: Intelligence from Tiny Substrates ​

Author: Giacomo Bocchese, Nicola Giacobbo, Etienne Guichard, James Wiles, Akshaj Devireddy
Published: 8/17/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2608.14019v1 Announce Type: cross Abstract: Emergent Models (EMs) are a machine learning paradigm based on simple yet open-ended substrates, such as cellular automata, in which modeling is treated not as the learning of a closed-form input-output map but as the emergence, within simple dynamic...

📖 Read original article


111. Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers ​

Author: Thiago Sandoval, Ufuk Topcu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CR, cs.LG

arXiv:2608.14089v1 Announce Type: cross Abstract: Safety classifiers deployed with large language models often fail for two reasons: their decisions reflect the policy learned during training rather than the deployer's desired policy, and their performance degrades as deployment traffic evolves. We ...

📖 Read original article


112. A Graph-Based Reinforcement Learning Framework for Structured Drift Diagnosis and Recovery in Autonomous LLM Agents ​

Author: Ismail El Hamraoui, Sagar Jose, Nicolas Bureau, Robert Plana
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2608.14109v1 Announce Type: cross Abstract: Autonomous LLM agents are increasingly deployed in complex real-world workflows, yet they remain vulnerable to runtime behavioral drift, a silent deviation from the original task that can lead to irreversible side effects on external systems. Existin...

📖 Read original article


113. Fixed-Budget Gaussian Volume Encoding with Structure-Aware Allocation ​

Author: Michael R. Martin, Joseph Insley, Victor A. Mateevitsi, Silvio Rizzi, Kwan-Liu Ma
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CE, cs.GR, cs.LG

arXiv:2608.14112v1 Announce Type: cross Abstract: Scientific simulations often produce scalar volumes faster than they can be stored, transferred, and loaded, while in situ reduction must use only a limited share of simulation resources. This work encodes scalar fields as anisotropic Gaussian primit...

📖 Read original article


114. AgilePE: Autonomous UAV Pursuit-Evasion via Self-Play Reinforcement Learning ​

Author: Wenhao Tang, Tianyang Chen, Zhejun Cui, Boyuan An, Jiayu Chen, Ruize Zhang, Huidong Liu, Tianyue Wu, Qingmin Liao, Fei Gao, Yu Wang, Chao Yu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.14135v1 Announce Type: cross Abstract: Autonomous pursuit-evasion is a fundamental challenge for Unmanned Aerial Vehicles (UAVs), requiring rapid decision-making under tightly coupled dynamics and continuously changing opponent behaviors. Traditional rule-based or differential-game approa...

📖 Read original article


115. Removing Temporal Note Redundancy Improves Multimodal Reinforcement Learning for Medicine ​

Author: Chenran Weng, Joo Seung Lee, Malini Mahendra, Anil Aswani
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.14157v1 Announce Type: cross Abstract: Mechanical ventilation is a critical life-support intervention, requiring dynamic adjustments to ventilator settings as a patient's condition evolves. While reinforcement learning (RL) offers a promising framework for optimizing these sequential deci...

📖 Read original article


116. Classical Limits of Spectral Filtering in Quantum Generative Models ​

Author: Marco Roth
Published: 8/17/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.14169v1 Announce Type: cross Abstract: Spectral filtering has been proposed as a route to regularization in quantum generative models: the quantum Fourier transform exposes the amplitude spectrum of a quantum circuit Born machine, and a diagonal filter suppresses the high frequencies asso...

📖 Read original article


117. Concept Guidance: Precise, Training-Free Latent Control for Text-to-Image Generation ​

Author: Nikolai R"ohrich, Isabell Hans, Felix Krause, Bj"orn Ommer
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.14172v1 Announce Type: cross Abstract: Text-to-image diffusion models have two major drawbacks that severely limit their practical utility: (1) standard models lack an intrinsic mechanism for continuous, concept-specific guidance (e.g., for precisely controlling how aesthetically pleasing...

📖 Read original article


118. FreeBalance: Pre-Routing Online Moe Load Balancing via Residual Workload Prediction ​

Author: Pengfei Chen, Yize Wu, Shouxu Kuang, Ke Gao, Ling Li
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.14205v1 Announce Type: cross Abstract: Load imbalance poses a major bottleneck to the efficiency of expert parallelism in distributed inference of Mixture-of-Experts (MoE) models. The most heavily loaded rank stalls global execution due to skewed routing distributions, directly increasing...

📖 Read original article


119. Attributing Preprocessing Invariance in Spectral Foundation Models ​

Author: Dongjun Wei, Hongyi Wu, Yinuo Zou
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CE, cs.LG

arXiv:2608.14227v1 Announce Type: cross Abstract: Preprocessing invariance is an appealing goal for spectral foundation models: a frozen model should remain useful when laboratories preprocess spectra differently. It is usually measured by training a classifier under one preprocessing pipeline and t...

📖 Read original article


120. Body size predicts how long ant workers live - but not how they age or how they die from heat ​

Author: Alana Moscardi, Rafael da Silva, Gleycon Silva
Published: 8/17/2026, 4:00:00 AM
Categories: q-bio.PE, cs.LG

arXiv:2608.14245v1 Announce Type: cross Abstract: In social insects, mortality risk comprises distinct components that may not share the same predictors: lifespan duration, senescence trajectory, and thermal vulnerability. We tested these three axes in 18 Australian ant species using paired field-la...

📖 Read original article


121. Meteorology-driven Causal Nowcasting of Fugitive Landfill Emissions Enables Proactive Public Health Response ​

Author: Timothy C. Pearce, David J. T. Smith, Alec Dobney, Alessia Freddo
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG, physics.ao-ph, physics.geo-ph

arXiv:2608.14254v1 Announce Type: cross Abstract: Fugitive emissions from waste sites increasingly expose communities to toxic and odorous gases, yet public-health responses remain largely retrospective, with episodes investigated only after residents have been exposed. Here we show that the meteoro...

📖 Read original article


122. Pairton: Iterative Reconstruction of Short-Lived Particles ​

Author: Andreas Hermansen, Chris Scheulen, Tobias Golling
Published: 8/17/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-ex

arXiv:2608.14278v1 Announce Type: cross Abstract: We present Pairton, an iterative framework for reconstructing short-lived particles in high-energy collision events. By formulating particle reconstruction as a masked prediction process over graph structures, Pairton learns conditional distributions...

📖 Read original article


123. Conditional Neural Optimal Transport for Predicting Cellular Phenotypes from Molecular Structure ​

Author: Gauthier Avit'e, Maxime Sanchez-Renauld, Nicolas Bourriez, Auguste Genovesio
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.14293v1 Announce Type: cross Abstract: High-content microscopy enables systematic profiling of cellular responses to chemical perturbations, but the scale of the chemical space makes exhaustive phenotypic characterization experimentally infeasible. This motivates computational models that...

📖 Read original article


124. Intelligent Detection of Mechanical, Electrical, and Plumbing (MEP) Metrics Based on 2D Floor Plans ​

Author: Tarandeep Singh Mandhiratta, ANK Zaman, Abdul-Rahman Mawlood-Yunis
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.HC, cs.LG

arXiv:2608.14317v1 Announce Type: cross Abstract: This research developed a neural network-based model to extract various information from 2D floor plans. We detect lighting symbols, identify the appropriate type of light, and extract the associated texts with lights. The study aims to enable effici...

📖 Read original article


125. A Four-Axis Trustworthiness Benchmark for LLM-as-Judge in Principle-Based Regulation ​

Author: Dipankar Sarkar
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.CY, cs.LG

arXiv:2608.14329v1 Announce Type: cross Abstract: Principle-based regulation, with evaluative standards such as "fair, clear, and not misleading" or "deliver good outcomes", cannot be reduced to binary predicates, and LLM-as-judge is increasingly used as the substitute. Our position is that any such...

📖 Read original article


126. CORAL: Curriculum-Optimized Reward Adaptation for LiDAR-Based Goal-Directed Urban Driving ​

Author: Anisa Saleem, Duksu Kim
Published: 8/17/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.14332v1 Announce Type: cross Abstract: Reinforcement learning is promising for autonomous urban driving, but long-horizon goal-directed navigation asks a policy to acquire several competing behaviors at once--reaching a distant goal, tracking a route, avoiding obstacles, obeying signals--...

📖 Read original article


127. Clearing the Fog: Towards Installing and Refining Proactive Exploration Capabilities in LLM Agents ​

Author: Zhizhao Guan, Chen Huang, Ziming Liu, Hongru Liang, Wenqiang Lei, See-Kiong Ng, Tat-Seng Chua, Anthony G Cohn
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.14339v1 Announce Type: cross Abstract: We study proactive exploration in LLM agents, i.e., the ability to explore an environment to acquire information that improves future decision-making. In this regard, we first identify two fundamental bottlenecks that hinder this capability and then ...

📖 Read original article


128. ATLAS: Discovering Agent Strategies through LLM-Guided Abstraction and Automata Learning ​

Author: Ignacio D. Lopez-Miguel, Andreas Happe, J"urgen Cito, Ezio Bartocci, Bettina K"onighofer, Martin Tappler
Published: 8/17/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2608.14352v1 Announce Type: cross Abstract: Large Language Model (LLM)-based agents are increasingly used for complex tasks such as software testing and cybersecurity assessment. While these agents demonstrate impressive capabilities, their behavior is difficult to understand, explain, and ana...

📖 Read original article


129. Non-Shattering at and Above the Dynamical Temperature in the Spherical Pure p-Spin Model ​

Author: Taegyun Kim
Published: 8/17/2026, 4:00:00 AM
Categories: math.PR, cond-mat.dis-nn, cs.LG, math-ph, math.MP

arXiv:2608.14369v1 Announce Type: cross Abstract: We consider the notion of shattering introduced by Ben Arous and Jagannath for spherical pure $p$-spin glasses with overlap $q$. For every $p\geq 3$ and $0\sqrt{(p-2)/(p-1)}$. The proof combines a deterministic $N+1$ bound for disjoint bands in the f...

📖 Read original article


130. Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages ​

Author: Chih-Hsuan Yang, Anjir Ahmed Chowdhury, Cheng-Hau Yang, Weijian Zheng, Fernando Llorente, Xiaolong Ma, Xinyang Li, Eliu A. Huerta, Ian T. Foster, Rajeev Thakur
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.14375v1 Announce Type: cross Abstract: Multi-agent reasoning systems often use agreement, confidence, or automated scores to decide which messages should shape a final answer. Such filtering assumes that a message likely to be correct is also worth keeping. Yet a wrong answer can contain ...

📖 Read original article


131. Offline Deep Q* Estimation with Diffusion Models ​

Author: Xiaohong Chen, Yuling Jiao, Lican Kang, Jerry Zhijian Yang, Chen Zhong
Published: 8/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.14401v1 Announce Type: cross Abstract: In offline RL, estimating the optimal action-value function $Q^*$ can be formulated as solving the optimal Bellman equation based solely on offline observations. A fundamental challenge is that the reward function and transition kernel are unknown, s...

📖 Read original article


132. Online Inference in Distributional Temporal-Difference Learning ​

Author: Yang Peng, Liangyu Zhang
Published: 8/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.14408v1 Announce Type: cross Abstract: We study online statistical inference for functionals of the return distribution under a fixed policy. The return distribution is estimated by nonparametric distributional temporal-difference learning from a single Markov trajectory. For the Polyak--...

📖 Read original article


133. Style or Signature? Artist-Disjoint Evaluation of Style Classification in Frozen Vision Embeddings ​

Author: Rory Ashton
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.14435v1 Announce Type: cross Abstract: Frozen image embeddings from models such as CLIP are increasingly used to classify paintings by art-historical style, with high reported accuracy. We ask whether this accuracy reflects an understanding of style or the recognition of individual artist...

📖 Read original article


134. You Only Pass Once: Answering and Abstaining Together in a Single Forward Pass of a Frozen Language Model ​

Author: Ziyang Luo, Zhongyao Chu, Xinjie He, Youting Wang, Xukui Qin, Runxiong Wu, Yan-Syuan Chen
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.14465v1 Announce Type: cross Abstract: A frozen language model on reasoning tasks has two coupled weaknesses: it under-uses evidence its own residual stream already encodes, and it fails to detect when the input is insufficient to answer, so it confabulates. This paper consolidates two re...

📖 Read original article


135. Expected Free Energy-based Informative Path Planning for Robotic Mars Exploration ​

Author: Ajith Anil Meera, Pablo Lanillos, Wouter Kouw
Published: 8/17/2026, 4:00:00 AM
Categories: cs.RO, cs.IT, cs.LG, math.IT

arXiv:2608.14466v1 Announce Type: cross Abstract: An autonomous robot efficiently exploring an unknown environment, such as looking for water sources on Mars, faces two simultaneous demands: building an accurate information map while quickly finding the regions of greatest value, and paying for ever...

📖 Read original article


136. Universal Thermodynamic Interatomic Potentials for Crystalline Materials ​

Author: Juno Nam, Bowen Deng, Xiaochen Du, Luis Barroso-Luque, Benjamin Kurt Miller, Rafael G'omez-Bombarelli
Published: 8/17/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cond-mat.stat-mech, cs.AI, cs.LG, physics.chem-ph

arXiv:2608.14502v1 Announce Type: cross Abstract: Free energies govern solid-state phase stability, yet computational materials discovery still relies largely on ground-state energies because free energy calculations require ensemble averages. We introduce the thermodynamic interatomic potential (TI...

📖 Read original article


137. Split the Labor: Separating Evidence Interpretation from Decision Aggregation ​

Author: Zhelun Wu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.14509v1 Announce Type: cross Abstract: Systems that ask a language model to reach a conclusion from many sources usually concatenate them into one prompt. This conflates two operations with different requirements. Interpreting a source rewards capacity and context. Combining interpretatio...

📖 Read original article


138. Decoding the Past: An Uncertainty-Aware Deep Learning Framework for Sex Attribution in Prehistoric Hand Stencils ​

Author: Karel Becerra, Boris Mederos, Dean Snow, Ram'on A. Mollineda
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.14539v1 Announce Type: cross Abstract: Determining the biological sex of the individuals who created Upper Paleolithic hand stencils remains a challenging problem due to the absence of ground truth, population differences between contemporary and prehistoric groups, and the uncertainty in...

📖 Read original article


139. Active Regression via Linear-Sample Sparsification ​

Author: Xue Chen, Eric Price
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.DS

arXiv:1711.10051v4 Announce Type: replace Abstract: We present an approach that improves the sample complexity for a variety of curve fitting problems, including active learning for linear regression, polynomial regression, and continuous sparse Fourier transforms. In the active linear regression pr...

📖 Read original article


140. Exposition on over-squashing problem on GNNs: Current Methods, Benchmarks and Challenges ​

Author: Dai Shi, Andi Han, Lequan Lin, Yi Guo, Junbin Gao
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2311.07073v3 Announce Type: replace Abstract: Graph-based message-passing neural networks (MPNNs) have achieved remarkable success in both node and graph-level learning tasks. However, several identified problems, including over-smoothing (OSM), limited expressive power, and over-squashing (OS...

📖 Read original article


141. A Probabilistic Framework for Learnable Optimization Algorithms ​

Author: Peter Ochs, Michael Sucker
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, math.PR

arXiv:2408.11629v2 Announce Type: replace Abstract: We propose a statistical-learning framework for optimization algorithms. The framework is based on probability distributions over optimization trajectories induced by a distribution of optimization problems and a learnable optimization algorithm. W...

📖 Read original article


142. OTIS: Learning High-Quality Time Series Features With Tiny Encoders ​

Author: "Ozg"un Turgut, Philip M"uller, Martin J. Menten, Daniel Rueckert
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2410.07299v3 Announce Type: replace Abstract: We introduce OTIS, an open time series encoder that yields high-quality time series features for downstream deployment on any system, including resource-constrained wearables and industrial sensors. Currently, the development of powerful general-pu...

📖 Read original article


143. Ordinal-Aware Calibration for Ordinal Classification ​

Author: Daehwan Kim, Haejun Chung, Ikbeom Jang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2410.15658v4 Announce Type: replace Abstract: Deep neural networks frequently produce overconfident, miscalibrated predictions. In ordinal classification, predictions must also adhere to a unimodal and order-consistent structure, a requirement that has dominated prior work while overlooking ca...

📖 Read original article


144. Generative Modeling with Bayesian Sample Inference ​

Author: Marten Lienen, Marcel Kollovieh, Stephan G"unnemann
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2502.07580v4 Announce Type: replace Abstract: We present a novel view of diffusion-like generative modeling from the perspective of iterative Gaussian posterior inference. By treating the generated sample as an unknown variable, we formulate the sampling process in the language of Bayesian pro...

📖 Read original article


145. Responsiveness Verification: Will Predictions Change? How Much? How Often? ​

Author: Harry Cheon, Meredith Stewart, Bogdan Kulynych, Tsui-Wei Weng, Berk Ustun
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2507.02169v2 Announce Type: replace Abstract: Machine learning models are often used in applications where their inputs change due to routine interactions, strategic manipulation, or noise. In such settings, models can undermine safety as these changes lead them to predict over regions of inpu...

📖 Read original article


146. Implicit Bias and Invariance: How Hopfield Networks Efficiently Learn Graph Orbits ​

Author: Michael Murray, Tenzin Chan, Kedar Karhadker, Christopher J. Hillar
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.14338v4 Announce Type: replace Abstract: Many learning problems are organized by group symmetries. While invariance is often imposed through architectures or group averaging, we ask when it can emerge from training on a finite random subset of an orbit. We study this question in classical...

📖 Read original article


147. INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT ​

Author: Idan Tankel, Nir Mazor, Rafi Brada, Christina LeBedis, Guy ben-Yosef
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, eess.IV

arXiv:2512.14732v3 Announce Type: replace Abstract: Incidental findings in CT scans, though often benign, can have significant clinical implications and should be reported following established guidelines. Traditional manual inspection by radiologists is time-consuming and variable. This paper propo...

📖 Read original article


148. PEFT-MuTS: A Multivariate Parameter-Efficient Fine-Tuning Framework for Remaining Useful Life Prediction based on Cross-domain Time Series Representation Model ​

Author: En Fu, Yanyan Hu, Zengwang Jin, Kaixiang Peng
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.22631v2 Announce Type: replace Abstract: The application of data-driven remaining useful life (RUL) prediction has long been constrained by the availability of large amount of degradation data. Mainstream solutions such as domain adaptation and meta-learning still rely on large amounts of...

📖 Read original article


149. ArGEnT: Arbitrary Geometry-encoded Transformer for Operator Learning ​

Author: Wenqian Chen, Zhi-Feng Wei, Yucheng Fu, Michael Penwarden, Pratanu Roy, Panos Stinis
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, physics.chem-ph, physics.comp-ph, physics.flu-dyn

arXiv:2602.11626v3 Announce Type: replace Abstract: Learning solution operators on arbitrary geometries remains a central challenge in scientific machine learning, especially for many-query simulation, physics-informed learning, and evolving geometries requiring accurate, geometry-aware predictions ...

📖 Read original article


150. Neural Network-Based Parameter Estimation of a Labour Market Agent-Based Model ​

Author: M Lopes Alves, Joel Dyer, Doyne Farmer, Michael Wooldridge, Anisoara Calinescu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.MA

arXiv:2602.15572v3 Announce Type: replace Abstract: Agent-based modelling (ABM) is a widespread approach to simulate complex systems. Advancements in computational processing and storage have facilitated the adoption of ABMs across many fields; however, ABMs face challenges that limit their use as d...

📖 Read original article


151. Global Interpretability via Automated Preprocessing: A Framework Inspired by Psychiatric Questionnaires ​

Author: Eric V. Strobl
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM, stat.ML

arXiv:2602.23459v2 Announce Type: replace Abstract: Psychiatric questionnaires are highly context sensitive and often only weakly predict subsequent symptom severity, which makes the prognostic relationship difficult to learn. Although flexible nonlinear models can improve predictive accuracy, their...

📖 Read original article


152. The Expressive Limits of Diagonal SSMs for State-Tracking ​

Author: Mehran Shakerinava, Behnoush Khavari, Siamak Ravanbakhsh, Sarath Chandar
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.01959v2 Announce Type: replace Abstract: State-Space Models (SSMs) have recently been shown to achieve strong empirical performance on a variety of long-range sequence modeling tasks while remaining efficient and highly-parallelizable. However, the theoretical understanding of their expre...

📖 Read original article


153. CarbonBench: A Global Benchmark for Upscaling of Carbon Fluxes Using Zero-Shot Learning ​

Author: Aleksei Rozanov, Arvind Renganathan, Yimeng Zhang, Vipin Kumar
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2603.09868v2 Announce Type: replace Abstract: Accurately quantifying terrestrial carbon exchange is essential for climate policy and carbon accounting, yet models must generalize to ecosystems underrepresented in sparse eddy covariance observations. Despite this challenge being a natural insta...

📖 Read original article


154. Exponential-Family Membership Inference: From LiRA and RMIA to BaVarIA ​

Author: Rickard Br"annvall
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2603.11799v2 Announce Type: replace Abstract: Membership inference attacks (MIAs) are becoming standard tools for auditing the privacy of machine learning models. The leading attacks -- LiRA (Carlini et al., 2022) and RMIA (Zarifzadeh et al., 2024) -- appear to use distinct scoring strategies,...

📖 Read original article


155. Optimization with SpotOptim ​

Author: Thomas Bartz-Beielstein
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.13672v2 Announce Type: replace Abstract: The spotoptim package implements surrogate-model-based optimization of expensive black-box functions in Python. Building on two decades of Sequential Parameter Optimization (SPO) methodology, it provides a Kriging-based optimization loop with Expec...

📖 Read original article


156. Untrained CNNs Match Backpropagation at V1: A Systematic RSA Comparison of Four Learning Rules Against Human fMRI ​

Author: Nils Leutenegger
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC

arXiv:2604.16875v3 Announce Type: replace Abstract: CORRECTION (August 2026): an evaluation-mode defect affected the predictive-coding and STDP conditions of this study; those results should not be used pending re-computation. At V1 and 224px, predictive coding falls from rho = 0.056 to 0.016 and ST...

📖 Read original article


157. Friction-Augmented Drifting Models for Resource-Efficient Domain Translation ​

Author: Arkadii Kazanskii, Tatiana Petrova, Andrey Ustyuzhanin, Konstantin Bagrianskii, Aleksandr Puzikov, Radu State
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2604.18194v2 Announce Type: replace Abstract: Single-step generators promise high-fidelity synthesis at a fraction of the inference and training cost of ordinary differential equation (ODE)-based flow models, a central concern when compute is limited. Drifting Models (DMs) train a one-step gen...

📖 Read original article


158. Inpainting physics: self-supervised learning for context-driven fluid simulation ​

Author: Jonas Weidner, Yeray Martin-Ruisanchez, Daniel Rueckert, Benedikt Wiestler, Julian Suk
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn

arXiv:2605.08832v3 Announce Type: replace Abstract: Neural surrogate models for computational fluid dynamics (CFD) are typically trained as forward operators that map explicit problem specifications, such as geometry and boundary conditions, to solution fields. This ties the model to the conditionin...

📖 Read original article


159. Cross-Species RSA Reveals Conserved Early Visual Alignment but Divergent Higher-Area Rankings Across Human fMRI and Macaque Electrophysiology ​

Author: Nils Leutenegger
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, q-bio.NC

arXiv:2605.22401v2 Announce Type: replace Abstract: CORRECTION (August 2026): an evaluation-mode defect in the shared feature-extraction pipeline affected the predictive-coding and STDP conditions. It applies to both sides of every comparison here: the human values are reprinted from the companion s...

📖 Read original article


160. Coupling-Robust Accuracy in Multiphysics Physics Informed Neural Networks via Kronecker-Preconditioned Optimization ​

Author: Youngjae Park, Jaemin Kim, Junghwa Hong
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2605.23391v3 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) offer a mesh-free route to solving coupled multiphysics systems, but their accuracy degrades systematically as inter-equation coupling strengthens, and inverse-gradient-norm loss balancing alone does not rel...

📖 Read original article


161. Localization then Neutralization: Gradient-guided Token Suppression against Visual Prompt Injection Attack ​

Author: Dongpeng Zhang, Ke Ma, Yangbangyan Jiang, Gaozheng Pei, Longtao Huang, Qianqian Xu, Qingming Huang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.25194v2 Announce Type: replace Abstract: Adversarial images pose a severe security threat to multimodal large language models through prompt injection. Existing defenses largely lack a principled understanding of the underlying mechanisms and struggle to balance efficiency and defense uti...

📖 Read original article


162. Supervised Training Rapidly Degrades Early Visual Cortex Alignment Across Biologically Plausible Learning Rules ​

Author: Nils Leutenegger
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC

arXiv:2605.30556v2 Announce Type: replace Abstract: CORRECTION (August 2026): the central finding of this paper is not supported. An evaluation-mode defect left the batch-normalisation layers of the predictive-coding and STDP conditions in training mode during feature extraction, producing their app...

📖 Read original article


163. Multimarginal flow matching with optimal transport potentials ​

Author: Raghav Kansal, David Crair, Nghia Nguyen, Scott Pope, Bradley Parry
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM, stat.ML

arXiv:2606.05327v2 Announce Type: replace Abstract: Flow matching (FM) has emerged as a powerful framework for learning dynamic transport maps between two empirical distributions. However, less explored is the setting with intermediate observed marginals that can help constrain the flows between the...

📖 Read original article


164. Learning Transfers: Kan Extensions for Neural Invariants ​

Author: Luciano Melodia
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, math.AT, math.CT

arXiv:2606.07627v3 Announce Type: replace Abstract: A representation transfers if it stays usable once the task has changed. Standard evaluations report target accuracy or a distance between data distributions, but neither says which structure of the representation is meant to survive. Here we make ...

📖 Read original article


165. Your Privacy My Cloak: Backdoor Attacks on Differentially Private Federated Learning ​

Author: Xiaolin Li, Ning Wang, Ninghui Li, Wenhai Sun
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2606.17035v2 Announce Type: replace Abstract: Prior research suggests that differential privacy (DP) inherently enhances the robustness of federated learning (FL) against backdoor attacks. In this paper, we challenge this assumption. Through an empirical analysis of two baseline attack strateg...

📖 Read original article


166. Breaking Chains with Trees: Model-Parallel Deep Learning with $\mathcal{O}(\log N)$ Time Complexity ​

Author: Neeraj Mohan Sushma, Aditya Nagarsekar, Cabrel Teguemne Fokam, Robin Schiewer, Amit Kumar Pal, Anand Subramoney, David Kappel
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DS

arXiv:2606.21497v2 Announce Type: replace Abstract: Modern deep neural networks are trained using error backpropagation, which requires sequential forward and backward computations across network layers. As these networks become deeper, this introduces limitations, since layer-wise updates are stric...

📖 Read original article


167. It's Much Easier for Neural Networks to learn Game of Life Dynamics with the Right Activation Function: Polynomial Kolmogorov-Arnold Networks ​

Author: Tashin Ahmed, Q. Tyrell Davis
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, nlin.CG

arXiv:2606.23587v2 Announce Type: replace Abstract: Previous work has found a gap between the scale of neural networks that reliably learn Conway's Game of Life, and minimal networks capable of representing the classic cellular automaton with hard-coded parameter values. Viewing neural network learn...

📖 Read original article


168. When Prices Double in a Week: Forecasting of Agricultural Volatility in Import-Isolated Markets ​

Author: Ranuga Weerasekara, Heshan Nethmina, Manuja Ranathunga, Vinma Wettasinghe, Dinithi Navodya, Subavarshana Arumugam, Nirasha Munasinghe, Nisansa de Silva, Sandareka Wickramanayake
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ME

arXiv:2606.29248v3 Announce Type: replace Abstract: Vegetable prices in Sri Lanka are highly volatile because the market is largely import-isolated, so supply disruptions quickly drive prices up. This study develops a machine learning framework to forecast such volatility by incorporating supply-cha...

📖 Read original article


169. Dissociating the Internal Representations of Sycophancy in LLMs ​

Author: Anthony Baez, Sheer Karny, Pat Pataranutaporn
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.07003v3 Announce Type: replace Abstract: Large Language Models (LLMs) frequently exhibit sycophancy, agreeing with a user's statement even when it is incorrect. While often studied as a single, uniform behavior, sycophancy can manifest in substantially distinct ways across contexts, raisi...

📖 Read original article


170. Neural Operator-enabled Topology-informed Evolutionary Strategy for PDE-Constrained Optimization ​

Author: Xiangming Huang (Georgia Institute of Technology), Guannan Zhang (Oak Ridge National Laboratory), Lu Lu (Yale University), Rapha"el Pestourie (Georgia Institute of Technology)
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.07682v2 Announce Type: replace Abstract: The inverse design of physical systems governed by partial differential equations is computationally demanding due to the high dimensionality and non-convexity of design spaces. Generative models for inverse design often lack robustness and transfe...

📖 Read original article


171. RUBRIC: Realism--Utility Balanced Ranking for Imbalanced Classification ​

Author: Yanxuan Yu, Dong Liu, Shu Wang, Wenxiao Zhao, Eric Jiang, Chang Liu, Jinxi Yu, Hui Pan, Ben Lengerich, Tong Geng, Renata Borovica-Gajic, Ying Nian Wu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09816v3 Announce Type: replace Abstract: Class imbalance poses a fundamental challenge in risk-sensitive applications such as fraud detection and medical diagnosis, where minority-class samples are scarce yet critical for accurate classification. Existing oversampling methods generate syn...

📖 Read original article


172. ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation ​

Author: Qingyu Zhang, Qianhao Yuan, Hongyu Lin, Yaojie Lu, Xianpei Han, Le Sun, Ming Xu, Jiarui Li
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.13124v2 Announce Type: replace Abstract: Structured pruning is a hardware-friendly way to compress LLMs, but it is mostly validated on multiple-choice recognition tasks, while the same compressed checkpoints can collapse on the free-form generation that deployment actually requires. Two o...

📖 Read original article


173. A Negative-Control Protocol for Clinical EEG Foundation-Model Benchmarks: Dataset Identity and External-Cohort Stress Testing ​

Author: Marzieh Zare
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE

arXiv:2607.24519v3 Announce Type: replace Abstract: EEG foundation-model gains may depend on cohort, montage, or probe design. We evaluated five models on five tasks across four benchmark datasets plus Korean CAUEEG, using subject-disjoint validation where identifiers exist. CAUEEG is recording-leve...

📖 Read original article


174. Subtype Robustness Is Not Just Accuracy: Calibration Under Unseen Subtype Shift ​

Author: Hanyu Su, Carlota Julbe i Juanola, Yibo Hu
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.00928v2 Announce Type: replace Abstract: Subtype robustness asks whether a model keeps the correct coarse prediction when test examples come from fine-grained subtypes absent from training but still inside a known coarse category. Prior work studies this almost entirely through accuracy. ...

📖 Read original article


175. Latent Reward Registers for Diffusion Preference Alignment ​

Author: Yuanshen Guan, Zipeng Feng, Chengru Song, Zhiwei Xiong, Peiqin Sun
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.03929v3 Announce Type: replace Abstract: Aligning diffusion models with human preferences usually relies on a sparse terminal reward evaluated on the final generated samples, which creates a severe temporal credit-assignment problem across the denoising process. We propose Latent Reward R...

📖 Read original article


176. From Non-Convex Self-Concordant Regularization to Scalable Quasi-Newton Training of PINNs ​

Author: Chenhao Si, Kang An, Shiqian Ma, Ming Yan
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.04206v2 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) often require high-accuracy quasi-Newton refinement to obtain reliable partial differential equation solutions, but their residual objectives can exhibit indefinite, nearly singular, and poorly scaled local ...

📖 Read original article


177. From Recoverability to Functional Use: Certifying Temporal Reports in Time-Series Forecasting ​

Author: Qipeng Qian, Yuntao Qian
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10433v3 Announce Type: replace Abstract: Models increasingly accompany time-series forecasts with temporal reports---delays, leading indicators, or selected history---yet a correct report need not describe the computation that produced the forecast. We formalize this as a three-stage cert...

📖 Read original article


178. Fast A/B/n Testing: Exact Multi-Policy Comparison via Tree-Coupled Feedback Sharing ​

Author: Yuxiao Wen
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.12831v2 Announce Type: replace Abstract: Online platforms increasingly compare many adaptive decision policies---ranking systems, recommendation algorithms, pricing rules, and language-model agents---while each reward-bearing interaction can be costly or risky. A direct A/B/n design gives...

📖 Read original article


179. Learning Discrete Decisions for MIPs with Constraint-Aware Diffusion ​

Author: Vincenzo Di Vito, Mehdi Taghizadeh, Deepjyoti Deka, Kaarthik Sundar, Ferdinando Fioretto
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.13079v2 Announce Type: replace Abstract: This paper proposes a novel learning-based approach to approximately solve instances of mixed-integer optimization problems. These problems are computationally challenging, as they require jointly determining discrete and continuous decisions while...

📖 Read original article


180. Simulation-to-real transfer learning for infrared spectroscopic chemical sensing and analysis from molecules to complex samples ​

Author: Yusen Tan, Yixuan Chen, Zheng Fang, Pan Liu, Yifan Li, Qinyu Guo, Zhedong Lin, Yuqiang Li, Xiangxiang Zeng, Tong Wang, Jun Xia
Published: 8/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.13341v2 Announce Type: replace Abstract: Infrared (IR) spectroscopy is widely used for chemical sensing, but extracting reliable chemical information from spectra remains challenging. Conventional interpretation is labor-intensive, relies on prior knowledge and reference spectra, and is d...

📖 Read original article


181. MLCC: A Congestion Control Technique to Accelerate ML Training ​

Author: Anton A. Zabreyko, Sanjoli Narang, Sudarsanan Rajasekaran, Manya Ghobadi
Published: 8/17/2026, 4:00:00 AM
Categories: cs.NI, cs.DC, cs.LG

arXiv:2402.09589v2 Announce Type: replace-cross Abstract: We present MLCC, a novel technique to augment today's congestion control algorithms to accelerate DNN training jobs in shared GPU clusters in a fully distributed manner. At the heart of MLCC lies a straightforward principle: DNN training flow...

📖 Read original article


182. Automated Inference of Graph Transformation Rules ​

Author: Jakob L. Andersen, Akbar Davoodi, Rolf Fagerberg, Christoph Flamm, Walter Fontana, Juri Kol\v{c}'ak, Christophe V. F. P. Laurent, Daniel Merkle, Nikolai N{\o}jgaard
Published: 8/17/2026, 4:00:00 AM
Categories: cs.DM, cs.LG, q-bio.MN

arXiv:2404.02692v3 Announce Type: replace-cross Abstract: The explosion of data available in life sciences is fueling an increasing demand for expressive models and computational methods. Graph transformation is a model for dynamic systems with a large variety of applications. We introduce a novel m...

📖 Read original article


183. Separation capacity of linear reservoirs with random connectivity matrix ​

Author: Youness Boutaib
Published: 8/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR

arXiv:2404.17429v4 Announce Type: replace-cross Abstract: A natural hypothesis for the success of reservoir computing in generic tasks is the ability of the untrained reservoir to map distinct input time series to separable reservoir states, a property we term separation capacity. In this work, we d...

📖 Read original article


184. READ: A Retrieval-Alignment Diffusion Framework for Structure-based Drug Design ​

Author: Dong Xu, Zhangfan Yang, Junchuang Cai, Sisi Yuan, Zexuan Zhu, Jianqiang Li, Junkai Ji
Published: 8/17/2026, 4:00:00 AM
Categories: q-bio.BM, cs.LG

arXiv:2506.14488v2 Announce Type: replace-cross Abstract: Structure-based drug design (SBDD) models are central to modern pharmaceutical research, enabling the rational exploration of protein-ligand interactions at atomic resolution. However, most existing approaches frame molecular generation as an...

📖 Read original article


185. PHASE: Passive Human Activity Simulation Evaluation ​

Author: Steven Lamp, Jason D. Hiser, Anh Nguyen-Tuong, Jack W. Davidson
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG, cs.NI

arXiv:2507.13505v2 Announce Type: replace-cross Abstract: Cybersecurity simulation environments, such as cyber ranges, honeypots, and sandboxes, require realistic human behavior to be effective, yet no quantitative method exists to assess the behavioral fidelity of synthetic user personas. This pape...

📖 Read original article


186. SALSA-V: Shortcut-Augmented Long-form Synchronized Audio from Videos ​

Author: Amir Dellali, Luca A. Lanzend"orfer, Florian Gr"otschla, Roger Wattenhofer
Published: 8/17/2026, 4:00:00 AM
Categories: cs.SD, cs.LG

arXiv:2510.02916v2 Announce Type: replace-cross Abstract: We propose SALSA-V, a multimodal video-to-audio generation model capable of synthesizing highly synchronized, high-fidelity long-form audio from silent video content. Our approach introduces a masked diffusion objective, enabling audio-condit...

📖 Read original article


187. A Configuration-First Framework for Reproducible, Low-Code Localization ​

Author: Tim Strnad (Jo\v{z}ef Stefan Institute, Slovenia), Bla\v{z} Bertalani\v{c} (Jo\v{z}ef Stefan Institute, Slovenia), Carolina Fortuna (Jo\v{z}ef Stefan Institute, Slovenia)
Published: 8/17/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2510.25692v4 Announce Type: replace-cross Abstract: As machine learning (ML) increasingly underpins critical applications, credible, comparable, and repeatable experimental results become more important. Everyday workflows should make rigorous experiment specification and controlled execution ...

📖 Read original article


188. The Nonstationarity-Complexity Tradeoff in Return Prediction ​

Author: Agostino Capponi, Chengpiao Huang, J. Antonio Sidaoui, Kaizheng Wang, Jiacheng Zou
Published: 8/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, q-fin.GN

arXiv:2512.23596v2 Announce Type: replace-cross Abstract: Does more data improve return prediction? In non-stationary financial markets, longer training windows improve prediction of complex models but incorporate outdated economic regimes, whereas simpler models require less data and are less vulne...

📖 Read original article


189. Semantic Differentiation for Tackling Challenges in Watermarking Low-Entropy Constrained Generation Outputs ​

Author: Nghia T. Le, Alan Ritter, Kartik Goyal
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2601.11629v2 Announce Type: replace-cross Abstract: We demonstrate that while the current approaches for language model watermarking are effective for open-ended generation, they are inadequate at watermarking LM outputs for constrained generation tasks with low-entropy output spaces. Therefor...

📖 Read original article


190. XtraLight-MedMamba for Classification of Neoplastic Tubular Adenomas ​

Author: Aqsa Sultana, Rayan Afsar, Ahmed Rahu, Surendra P. Singh, Brian Shula, Brandon Combs, Derrick Forchetti, Vijayan K. Asari
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2602.04819v5 Announce Type: replace-cross Abstract: Accurate risk stratification of precancerous polyps during routine colonoscopy screening is a key strategy to reduce the incidence of colorectal cancer (CRC). However, assessment of low-grade dysplasia remains limited by subjective histopatho...

📖 Read original article


191. Semantically Labelled Automata for Multi-Task Reinforcement Learning with LTL Instructions ​

Author: Alessandro Abate, Giuseppe De Giacomo, Mathias Jackermeier, Jan Kret'insk'y, Maximilian Prokop, Christoph Weinhuber
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2602.06746v2 Announce Type: replace-cross Abstract: We study multi-task reinforcement learning (RL), a setting in which an agent learns a single, universal policy capable of generalising to arbitrary, possibly unseen tasks. We consider tasks specified as linear temporal logic (LTL) formulae, w...

📖 Read original article


192. A Systematic Comparison of Training Objectives for Out-of-Distribution Detection in Image Classification ​

Author: Furkan Gen\c{c}, Onat "Ozdemir, Emre Akba\c{s}
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2603.07571v3 Announce Type: replace-cross Abstract: Out-of-distribution (OOD) detection is critical in safety-sensitive applications. While this challenge has been addressed from various perspectives, the influence of training objectives on OOD behavior remains comparatively underexplored. In ...

📖 Read original article


193. Early Stopping for Large Reasoning Models via Confidence Dynamics ​

Author: Parsa Hosseini, Sumit Nawathe, Mahdi Salmani, Meisam Razaviyayn, Soheil Feizi
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.04930v2 Announce Type: replace-cross Abstract: Large reasoning models rely on long chain-of-thought generation to solve complex problems, but extended reasoning often incurs substantial computational cost and can even degrade performance due to overthinking. A key challenge is determining...

📖 Read original article


194. Retrieve-then-Adapt: Retrieval-Augmented Test-Time Adaptation for Sequential Recommendation ​

Author: Xing Tang, Ziqiang Cui, Jingyang Bin, Xiaokun Zhang, Fuyuan Lyu, Jingyan Jiang, Dugang Liu, Chen Ma, Xiuqiang He
Published: 8/17/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2604.05379v2 Announce Type: replace-cross Abstract: The sequential recommendation (SR) task aims to predict the next item based on users' historical interaction sequences. Typically trained on historical data, SR models often struggle to adapt to real-time preference shifts during inference du...

📖 Read original article


195. Learning-Guided Sparsification of Dynamic Graphs in Robotic Exploration ​

Author: Adithya V. Sastry, Bibek Poudel, Weizi Li
Published: 8/17/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2604.16509v2 Announce Type: replace-cross Abstract: Many robotic exploration algorithms rely on graph structures for frontier-based exploration and dynamic path planning. However, these graphs grow rapidly, accumulating redundant information and impacting performance. We present a hybrid trans...

📖 Read original article


196. OlmoEarth v1.2: A more efficient family of OlmoEarth models ​

Author: Gabriel Tseng, Yawen Zhang, Favyen Bastani, Henry Herzog, Joseph Redmon, Hadrien Sablon, Piper Wolters, Ando Shah, Patrick Alan Johnson, Christopher Wilhelm, Patrick Beukema
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2605.20804v3 Announce Type: replace-cross Abstract: We present a set of improvements to the OlmoEarth family. These improvements allow us to cut compute costs during training ($3.0 \times$ reduction in GPU hours required to train our Base models) and inference ($2.9\times$ reductions in MACs o...

📖 Read original article


197. PhoneWorld: Scaling Phone-Use Agent Environments ​

Author: Yuxuan Liu, Xin Lai, Junyi Li, Pengyuan Lyu, Jason, Yiduo Guo, Zhengyao Fang, Yang Ding, Yi Zhang, Weinong Wang, Huawen Shen, Xingran Zhou, Liang Wu, Fei Tang, Sunqi Fan, Shangpin Peng, Zheng Ruan, Anran Zhang, Chengquan Zhang, Han Hu, Benyou Wang, Ji-Rong Wen, Rui Yan, Zhengyang Tang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2605.29486v2 Announce Type: replace-cross Abstract: A central bottleneck for phone-use agents is that controllable, reproducible environments covering real mobile behavior are hard to build at scale. Existing mobile-agent benchmarks have made important progress on evaluation, but they do not b...

📖 Read original article


198. TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration ​

Author: Soyeong Jeong, Jinheon Baek, Minki Kang, Sung Ju Hwang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2606.04743v2 Announce Type: replace-cross Abstract: Agents are widely deployed as assistants over documents, tools, and code. However, they typically act only on explicit user requests, which surface only the problems the user has noticed, while many other important problems coexist, hidden in...

📖 Read original article


199. Revisiting the shutdown problem ​

Author: David Thorstad
Published: 8/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.08296v2 Announce Type: replace-cross Abstract: A key premise in leading arguments for existential risk from artificial intelligence is that malfunctioning artificial agents could not be easily shut down. This motivates the catastrophic shutdown problem of ensuring that agents can be shut ...

📖 Read original article


200. OCOO-T : A Simple and Scalable Virtual Cell Model for Transcriptional Perturbation Response Prediction ​

Author: Danning Jiang, Zhiwen Yan, Qirun Wang, Zheming An, Yalong Zhao, Lipeng Lai
Published: 8/17/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG, q-bio.GN

arXiv:2606.12838v2 Announce Type: replace-cross Abstract: Predicting single-cell transcriptional responses to genetic, chemical and cytokine perturbations is a fundamental challenge in computational biology and AI Virtual Cell (AIVC) modeling, with direct implications for drug discovery and the eluc...

📖 Read original article


201. RL-Index: Reinforcement Learning for Retrieval Index Reasoning ​

Author: Yongjia Lei, Nedim Lipka, Zhisheng Qi, Utkarsh Sahu, Yuchen Zhuang, Wenqi Shi, Koustava Goswami, Franck Dernoncourt, Ryan A. Rossi, Yu Wang
Published: 8/17/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2606.16316v2 Announce Type: replace-cross Abstract: Retrieving external knowledge is crucial for real-world tasks but remains difficult when queries and relevant knowledge are linked by implicit reasoning (e.g., shared theorems or coding logic). Existing methods rely mainly on query-side reaso...

📖 Read original article


202. Point-Cloud-Assistant Localized Statistical Channel Prediction by Tangent Gaussian Splatting ​

Author: Ye Xue, Yiheng Wang, Xinhua Shao, Qi Yan, Shutao Zhang, Tsung-Hui Chang
Published: 8/17/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2606.18734v2 Announce Type: replace-cross Abstract: Accurate, site-specific channel information is crucial for optimizing next-generation wireless networks. Among various approaches, localized statistical channel modeling (LSCM), which models the channel multipath angular power spectrum (APS) ...

📖 Read original article


203. Cross-Calibrated Confidence Fields for Local Risk Updates ​

Author: Mingzhi Song
Published: 8/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2606.19147v3 Announce Type: replace-cross Abstract: How can training data be used to compare local updates to the current model, choose an update, and retain valid bounds for the selected update's population-risk change? We construct lower and upper confidence fields that jointly cover the pop...

📖 Read original article


204. Event-Conditioned Diagnostics of Kinematic, Contact, and Object-Permanence Fields in Passive Object-State World Models ​

Author: Yang Liu, Yuming Chen
Published: 8/17/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2606.28455v2 Announce Type: replace-cross Abstract: World models can predict future physical states, but prediction accuracy alone does not explain how physical information is organized and used inside their latent dynamics. We introduce a controlled diagnostic protocol for studying event-cond...

📖 Read original article


205. $\mu$Flow: Leveraging Average Images for Improving Generalisation of Deepfake Faces Detectors ​

Author: Orazio Pontorno, Mattia Litrico, Luca Guarnera, Mario Valerio Giuffrida, Sebastiano Battiato
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2606.30528v2 Announce Type: replace-cross Abstract: Current generative models, including GANs and diffusion models, have reached an outstanding level of photorealism, posing significant risks to privacy and security. To ensure real-world applicability, deepfake detectors must generalise effect...

📖 Read original article


206. Debiasing Text-to-Image Evaluation via Implicit Cultural Alignment Reward Modeling ​

Author: Bo-An Chang, Yu-Chih Chen
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MM

arXiv:2607.15740v2 Announce Type: replace-cross Abstract: As Text-to-Image (T2I) systems rapidly advance, evaluating the cultural authenticity of synthesized content has become increasingly important for fair and trustworthy generative AI. Existing T2I evaluation metrics and multimodal judges often ...

📖 Read original article


207. A Direct Route to Markov Chain Convergence via Asymptotic Equivalence with the Target ​

Author: Patrick Forr'e
Published: 8/17/2026, 4:00:00 AM
Categories: math.PR, cs.LG, math.ST, stat.CO, stat.ML, stat.TH

arXiv:2608.03353v2 Announce Type: replace-cross Abstract: For a Markov kernel $T$ with an invariant probability measure $\pi$, we give a self-contained proof of the Markov chain convergence theorem via a criterion called asymptotic equivalence with the target. It assumes two parts about the Lebesgue...

📖 Read original article


208. Distribution-Free Conformal Prediction for Steel Fatigue Strength: Marginal Validity Is Not Enough ​

Author: Irene Boruah
Published: 8/17/2026, 4:00:00 AM
Categories: stat.AP, cs.LG

arXiv:2608.07589v2 Announce Type: replace-cross Abstract: Predicting fatigue failure in steel components experimentally is costly, requiring testing across multiple compositions and processing conditions, spurring research on data-driven prediction models. Studies using the NIMS MatNavi steel fatigu...

📖 Read original article


209. From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability ​

Author: Alexander Hackett, Arnaud Denis-Remillard, Axel Cassou
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.08904v2 Announce Type: replace-cross Abstract: How much of a vision-language model's (VLM) spatial understanding remains after the action post-training process of building a vision-language-action model (VLA)? We probe depth perception, a primitive of spatiogeometric understanding, from e...

📖 Read original article


210. Test-Time Scaling for CAD Generation via Verifier-Free Consensus Selection ​

Author: Aaron Haag, Altay Kacan, Bertram Fuchs, Oliver Lohse
Published: 8/17/2026, 4:00:00 AM
Categories: cs.CE, cs.LG

arXiv:2608.09706v2 Announce Type: replace-cross Abstract: Large language models can write parametric CAD programs from a natural-language description (text-to-CAD generation), but a single sample is often wrong. Increasing test-time compute by sampling multiple candidates only helps if a good candid...

📖 Read original article


211. AI-Driven Multiscenario Interest Rate Forecasting: A Proof of Concept for Banking Asset Management ​

Author: Ekkehardt Bauer, Dirk Holl"ander, David Scholz, Linus Wolff, Christoph Ostermair, Kyrillus Aiad, Joachim Hasebrook
Published: 8/17/2026, 4:00:00 AM
Categories: q-fin.CP, cs.LG, q-fin.ST

arXiv:2608.12424v2 Announce Type: replace-cross Abstract: This study focuses on developing an AI-supported prototype for multiperspective interest rate forecasting that combines classical econometric models with modern artificial intel-ligence methods. Tested in a major European bank, the system ena...

📖 Read original article


212. Online Inference for Quantile Temporal Difference Learning in Distributional Reinforcement Learning ​

Author: Zijie Cheng, Yang Peng, Zhihua Zhang
Published: 8/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.12973v2 Announce Type: replace-cross Abstract: In this paper, we study how to perform statistical inference for quantile temporal difference learning (QTD) in distributional reinforcement learning. Assuming access to a generative model, we first establish functional central limit theorems...

📖 Read original article