arXiv cs.LG - 2026-09-02 ​
323 items collected.
1. Task-Specific Prompt with Global Context for Multi-Task Graph Pre-Training ​
Author: Zhiyang Qiu, Yangtao Wang, Xiaocui Li, Yanzhao Xie, Siyuan Chen, Wensheng Zhang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00047v1 Announce Type: new Abstract: Graph prompt learning is an effective paradigm to adapt pre-trained graph models to downstream tasks in low-resource scenarios. However, existing multi-task graph pre-training frameworks generally use randomly initialized prompts, leading to poor align...
2. REAL-Q: E2E LLM Quantization via Dynamic Gradient Descent ​
Author: Qian Zhang, Yaoming Li, Zhewen Tan, Yanshu Wang, Heng Lu, Kun Su, Zongwei Lv, Wenhan Yu, Yongge Ma, Yinjun Han, Ruikuang Liu, Tong Yang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00049v1 Announce Type: new Abstract: Post-training quantization (PTQ) is essential for deploying large language models (LLMs) under strict resource constraints. State-of-the-art PTQ methods quantize each layer with a single closed-form second-order solver: to remain analytically tractable...
3. Convergence issues in Relational Concept Analysis based on AOC-posets ​
Author: Xavier Dolques, Agn`es Braud, Alain Gutierrez, Marianne Huchard, Florence Le Ber
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00054v1 Announce Type: new Abstract: Formal Concept Analysis (FCA) is an approach for conceptual classification building and rule discovery from a binary table describing a set of objects by a set of attributes. Extensions have been proposed to deal with non-binary and more complex data, ...
4. DISTAL: Distillation and Self-Supervised Pretraining for Structure-Agnostic Materials Property Prediction ​
Author: Weiran Wang, Xintong Huo, Yueying Wang, Yusi Fan, Wenyan Wang, Xin Feng, Ruihao Xin, Lan Huang, Kewei Li, Fengfeng Zhou
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET
arXiv:2609.00059v1 Announce Type: new Abstract: Materials property prediction remains difficult in low-data settings, where many target properties are supported by only a limited number of labeled samples. Models with the strongest predictive accuracy often depend on crystal structures, which restri...
5. ReNFT: Repairing Mode Collapse in Reward Post-Training via Internal Probability-Mass Recalibration ​
Author: Yuchen Bao, Chao Wen, Haowei Wang, Ruoxin Chen, Donghao Luo, Jiahui Zhan, Wenjian Huang, Shen Chen, Yiting Wang, Taiping Yao, Chengjie Wang, Shouhong Ding, Jianguo Zhang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2609.00061v1 Announce Type: new Abstract: Reward post-training of diffusion generators inevitably concentrates probability mass on a few reward-favored modes, a mode collapse that erases within-prompt diversity. Existing methods for mitigating collapse rely on external signals or interfaces, a...
6. Attention Sensitivity Is Not Enough: Dissociating Attention-Level and Behavioural In-Context Learning under Fine-Tuning ​
Author: Jinyuan Zhang, Peng He, He Hu, Yin Yuan, ShengShuo Jiao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2609.00064v1 Announce Type: new Abstract: In-context learning (ICL) lets large language models adapt to new tasks from demonstrations, and fine-tuning can erode this behaviour. Many preservation diagnostics inspect attention: if attention changes when demonstrations change, the model is treate...
7. RW-LoRA: Communication-Efficient Decentralized LoRA Fine-Tuning via Random Walks ​
Author: Xingran Chen, Rohit Bhagat, Ghadir Ayache, Rawad Bitar, Yanmin Gong, Salim El Rouayheb
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00078v1 Announce Type: new Abstract: Parameter-efficient fine-tuning methods such as LoRA have become a standard approach for adapting large foundation models. Adopting fine-tuning to distributed settings faces several challenges. Most existing distributed LoRA methods rely on centralized...
8. Stochastic complexity of vectors containing cluster structure ​
Author: Daniel Nicorici, Olli Yli-Harja, Jaakko Astola
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, stat.ML
arXiv:2609.00084v1 Announce Type: new Abstract: This paper studies the problem of computing the stochastic probability (shortest code length) of the encoded vectors containing cluster structure using Normalized Maximum Likelihood (NML) model. This is of great theoretical and practical importance in ...
9. Foundation models for electricity price forecasting and battery arbitrage: Can they replace market-specific forecasting models? ​
Author: Arkadiusz Lipiecki, Rafa{\l} Weron
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, econ.EM
arXiv:2609.00089v1 Announce Type: new Abstract: Foundation models promise accurate forecasts with little or no task-specific training, but whether they can replace models designed specifically for electricity price forecasting remains unclear. We compare nine variants from five foundation model fami...
10. Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence ​
Author: Eddie Conti, Claudio Daka, 'Alvaro Parafita, Antonio L. Alfeo, Axel Brando, Mario G. C. A. Cimino
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00090v1 Announce Type: new Abstract: Feature importance Methods (FIMs) are widely used in Explainable AI to interpret model predictions, yet attribution scores alone often provide limited insight into the underlying reasoning process. In this work, we introduce a novel perspective by embe...
11. Safin-1: Safety from Within through Memory-Native State Evolution ​
Author: Ming Zhang, Kaisen Yang, Shu Yu, Ermo Hua, Zhekai Chen, Cheng Jin, Jingnan Zheng, Yi Zhang, Zhongtian Ma, Jiawei Zhou, Sirui Chen, Qiaosheng Zhang, Xiang Wang, Ning Ding, Xia Hu, Bowen Zhou, Youbang Sun, Chaochao Lu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00092v1 Announce Type: new Abstract: Long-horizon complex tasks require foundation models to accumulate information, maintain internal states, and adapt over extended interactions. Safety should be an intrinsic property of the model itself, rather than a behavioral constraint relying sole...
12. Local Reference Geometry Residual Augmentation for Imbalanced Time Series Classification ​
Author: Chuanhang Qiu, Yanran Xu, Yue Wang, Anthony Bagnall
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00093v1 Announce Type: new Abstract: Imbalanced time series classification is often addressed by changing the training distribution, objective, logits, or final threshold. These interventions address important biases, yet leave a representation-level question unmeasured: after minority su...
13. Faster Than Flash: Exploiting Attention Sparsity for Efficient Long-Context Decoding ​
Author: Zhigeng Liu, Zhiyuan Ning, Ruixiao Li, Xiaoran Liu, Yuerong Song, Min Zhang, Ziwei He, Xipeng Qiu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00097v1 Announce Type: new Abstract: The development of long-context Large Language Models (LLMs) is constrained by the memory bandwidth bottleneck and quadratic complexity of the attention mechanism during decoding. To overcome the inherent trade-offs between the memory overhead of metad...
14. Generative artificial intelligence for reliable mechanistic reasoning for corrosion ​
Author: Bharath M N, R K Singh Raman, Alankar Alankar
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci
arXiv:2609.00099v1 Announce Type: new Abstract: Corrosion accounts for approximately 4% of global GDP, and reliable prediction is essential for timely mitigation. Machine learning effectively predicts corrosion rates from composition, microstructure, and environmental variables, but cannot explain t...
15. Good Memory Has ECC: Evaluating the Memory of Vision-Language Models Beyond Accuracy ​
Author: Shmuel Berman, Jia Deng
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00103v1 Announce Type: new Abstract: Memory is widely viewed as an important unsolved problem for LLMs and VLMs, and current benchmarks typically evaluate it by testing accuracy over long text or video. However, accuracy alone misses properties that matter for real long-horizon tasks. We ...
16. Flawed in Nature, Perfect through Evolution ​
Author: J. M. Diederik Kruijssen (Allora Foundation)
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2609.00129v1 Announce Type: new Abstract: The performance of artificial intelligence (AI) and machine learning (ML) models degrades when the problem they were trained on drifts. This is a near-universal feature of real-world problems, which often change unpredictably. Biological evolution has ...
17. Elite-Weighted Supervised Fine-tuning for Goal-Directed Molecular Optimization ​
Author: Shiyun Wa, Yifei Wang, Anna G. Green, Simone Sciabola, Ye Wang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00189v1 Announce Type: new Abstract: Goal-directed optimization is essential for steering molecular generators to propose candidates with desired properties. However, it is often implemented with policy-gradient reinforcement learning, which requires a generation-trajectory log-probabilit...
18. WHALE: A Simple Recipe for Joint Harness-Weight Optimization ​
Author: Haechan Kim, Yoonho Lee, Gisang Lee, Chelsea Finn, Kangwook Lee
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00196v1 Announce Type: new Abstract: Agent performance depends jointly on the model parameters and the executable harness code that manages context and control flow. Optimizing either component in isolation can leave the system bottlenecked by its frozen counterpart: weight updates can ch...
19. QTEA: Ternary LLMs with Sparse Residual Salient Weight and By-Column Optimization ​
Author: Yipin Guo, Arun M George, Jie Fu, Tareq Mahmoud, Sixue Xing, Siddharth Joshi
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00224v2 Announce Type: new Abstract: Weight-only post-training quantization (PTQ) can alleviate the computational burden of serving large language models (LLMs) at scale. However, existing PTQ methods often fail to generalize across models and suffer severe accuracy loss below 2 bits. Man...
20. Geometry-aware Latent Autoregressive Generative Model for PDEs in Complex Domains ​
Author: Zi Wang, Minghui Xu, Tapan Mukerji
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00297v1 Announce Type: new Abstract: Solving multiphysics partial differential equations (PDEs) remains a major challenge in scientific computing, especially for highly complex $\mu$m-scale tortuous geometries critical to energy and chemical engineering. We address this challenge by propo...
21. Do LLMs Know Your Neighborhood? Auditing LLM Priors for Neighborhood-Level Mobility Prediction and Structural Alignment ​
Author: Saad Mohammad Abrar, Eesha Kurella, Arnav Dadarya, Naman Awasthi, Kazi Tasnim Zinat, Vanessa Frias-Martinez
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CY
arXiv:2609.00345v1 Announce Type: new Abstract: Human mobility is central to urban planning, transportation, public health, and emergency response, yet fine-grained trajectory data are often proprietary, restricted, and privacy-sensitive. Large language models (LLMs) offer a potential alternative by...
22. Deterministic LLM Inference Across GPU Kernels: Power-of-Two INT8 Quantization Scales and the Limits of Tolerance-Based Conformance ​
Author: Teng-Ruei Chen
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.SE
arXiv:2609.00363v1 Announce Type: new Abstract: Conformance suites for quantized GEMM kernels ask whether two implementations agree within a tolerance. We measure what such a suite can detect. Injecting nine faults into a reference INT8 pipeline over 8,232 layer--fault--regime cells of Qwen3-1.7B, w...
23. Counterfactual Fragility Certificates: Exposing High-Confidence Brittleness under Structured Evidence Failure ​
Author: Filippo Cenacchi, Longbing Cao, Runze Yang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00366v1 Announce Type: new Abstract: High test accuracy and good aggregate calibration do not show whether an individual prediction is structurally supported by its evidence. In tabular decision systems, failures often occur when a feature family becomes unavailable, delayed, noisy, stale...
24. Adapting Without Gradients: Affine Statistics Transport and What Its Certificate Can Tell You ​
Author: Salim Khazem, Ibrahim Mohamed Serouis
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2609.00374v1 Announce Type: new Abstract: Test-time adaptation (TTA) typically assumes that model parameters can be updated at inference time. This assumption is restrictive for inference-only accelerators, frozen or third-party models, and memory-constrained deployments, and standard BatchNor...
25. Neural means and kernel corrections for operator learning ​
Author: Yitzchak Shmalo
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, math.PR, stat.ML
arXiv:2609.00389v1 Announce Type: new Abstract: We combine neural network means with exact Mat'ern kernel regressions of their residuals and of their learned features, and evaluate the pairing on two public emulation problems with published baselines: the structural-mechanics benchmark of de Hoop e...
26. A Multi-Branch Feature Fusion Approach for Health Misinformation Detection and Propagation ​
Author: Mkululi Sikosana, Sean Maudsley-Barton, Oluwaseun Ajao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2609.00403v1 Announce Type: new Abstract: This paper presents a multi-branch fusion framework for detecting and characterising the propagation of health misinformation in online social networks (OSNs). Grounded in the Elaboration Likelihood Model (ELM) and the Theory of Planned Behaviour (TPB)...
27. How Temporal Correlations Shape Memory in Linear Recurrent Neural Networks ​
Author: Arnol Manuel Fokam, Fasseu Sieyondji Akpevwoghene, Edem Fiifi Dawson
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00420v1 Announce Type: new Abstract: The linear recurrent neural network (LRNN) is a simple model for studying how much memory a network builds up as it trains. For uncorrelated inputs, earlier work found that training itself settles the network between keeping the past and reacting only ...
28. Group Adaptive Clipping Policy Optimization ​
Author: Sheng Jia, Xiao Wang, Shiva Prasad Kasiviswanathan, Rein Houthooft
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2609.00444v1 Announce Type: new Abstract: Group relative policy optimization for reinforcement learning with verifiable rewards (RLVR) typically uses a fixed importance-sampling (IS) ratio clipping boundary across all rollouts. We identify a key limitation: rare correct rollouts on harder prob...
29. CRAD: Class-wise Reliability-Aware Distillation for Decentralized Heterogeneous Federated Learning ​
Author: Baraa Bilbeisi, Mengchen Fan, Baocheng Geng, Qing Tian
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2609.00446v1 Announce Type: new Abstract: Conventional federated learning (FL) relies on parameter averaging, which forces clients to be doubly homogeneous: it demands an identical architecture and degrades under non-IID data. Real-world deployments usually break both assumptions. We sidestep ...
30. HBQ: Hierarchical Scaling Block Quantization with Hardware-Efficiency-Aware Design for Accurate LLM Inference ​
Author: Chun-Ting Chen, Dongmin Han, Hangyeol Mun, Jake Hyun, Arnab Raha, Amit Agarwal, Mark Anders, Mohamed Abdelfattah, Jae-sun Seo
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR
arXiv:2609.00450v1 Announce Type: new Abstract: Block Quantization (BQ) is a promising approach for efficient deployment of large language models (LLMs), enabling low-precision computation with controlled accuracy degradation. Compared to scalar weight-only quantization (WoQ), BQ quantizes both weig...
31. Can LLMs Use Relational Transformer Embeddings? ​
Author: Francisco Galuppo Azevedo, Clarissa Lima Loures
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00457v1 Announce Type: new Abstract: Injecting frozen relational-encoder embeddings as soft tokens into a large language model (LLM) is a conceptually appealing fusion strategy: the encoder handles multi-table structure, the LLM handles language and reasoning, and no lossy text serializat...
32. Context Window Failures in Relational Foundation Models ​
Author: Denis Oliveira Correa, Francisco Galuppo Azevedo
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00460v1 Announce Type: new Abstract: Recent Relational Deep Learning architectures have been proposed as foundation models for multi-table relational data, yet they impose constrained neighborhood budgets that force row truncation when an entity has many related records. We introduce Anim...
33. Higher Structures in Deep Learning ​
Author: Michael L. Roberts, Carlos Zapata Carratal'a. Nicholas J. Cooper, Lijun Chen, Fran\c{c}ois G. Meyer, Danna Gurari
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00472v1 Announce Type: new Abstract: We provide an expository introduction on the importance of higher-arity tensor operations to deep learning. Then, we conduct a novel empirical investigation of higher-arity phenomenon in trained neural networks, introduce a hypergraphical generalizatio...
34. AdaptNTK: Adaptive Uncertainty Quantification and Active Learning for Neural Network Potentials ​
Author: Prajwal Ananth, Shuwen Yue
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph
arXiv:2609.00488v1 Announce Type: new Abstract: Machine learning interatomic potentials bridge the gap between quantum chemical precision and classical computational speed, enabling molecular dynamics simulations with first-principles accuracy. Their reliability is often improved through active lear...
35. A hybrid quantum-classical neural network for learning to route ​
Author: Marcus Rolf Peter Ritt, Alexsandro Santos da Rosa J'unior, Marcos Vinicius Reballo, Cesar Augusto do Amaral, Fernando Augusto Caletti de Barros
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, quant-ph
arXiv:2609.00489v1 Announce Type: new Abstract: This work studies hybrid quantum-classical neural networks for learning routing heuristics. Specifically, this paper asks whether small quantum neural networks can replace parameter-heavy modules inside a competitive attention-based routing model while...
36. VATO: A Vortex-Force-Aware Transformer Operator for Unsteady Separated Aerofoil Flows ​
Author: Xingxin Yang, Zhan Zhang, Yichen Li, Juan Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00507v1 Announce Type: new Abstract: Accurate prediction of unsteady separated flows is challenging because the aerodynamic loads depend on nonlinear separation and vortex-shedding dynamics. Although high-fidelity CFD resolves these mechanisms, its cost limits repeated use in design and c...
37. Learning Task-Specific Antibody Representations via Function-Aware Masking ​
Author: Ayan Goel, Thomas A. Walton, Amirali Aghazadeh
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM
arXiv:2609.00518v1 Announce Type: new Abstract: Antibody-specific language models pretrained via masked language modeling (MLM) learn representations that are critical for downstream sequence design and property prediction tasks. Yet, the corruption process itself is rarely leveraged as a source of ...
38. Why Multi-Layer Message Passing Works: Completeness Theory for Graph Neural Network Interatomic Potentials ​
Author: Pingbing Ming, Han Wang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, physics.chem-ph, physics.comp-ph
arXiv:2609.00528v2 Announce Type: new Abstract: We prove that the Hypergraph Neural Network, an invariant architecture with 3-body message passing, is a universal approximator for potential energy surfaces. Our main contribution is a multi-layer completeness theory. We show that $L$ layers of messag...
39. DeSyR: A Decoupled Symbolic Recovery Framework with PINN-Guided Structure Search and Physics-Informed Coefficient Refinement ​
Author: Pancheng Niu, Jun Guo, Qiaolin He, Jingcai Guo, Yanchao Shi
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00530v1 Announce Type: new Abstract: Recovering compact explicit solutions from neural approximations is challenging when imperfect teacher data guide symbolic topology search and coefficient estimation. We present DeSyR, a decoupled symbolic recovery framework for differential equations....
40. GenONet: A Generative operator Network for High-Resolution Precipitation Nowcasting ​
Author: Mohammad Kian Golkar, Luciano Alves de Oliveira, Mohammad Khanjani
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph
arXiv:2609.00544v1 Announce Type: new Abstract: High-resolution precipitation nowcasting is critical for reducing the impacts of severe weather but remains difficult because of rapid storm evolution. Deep learning models have shown great promise for this task, but their predictive skill often deteri...
41. Manifold-Aware General Coded Computing for Straggler-Resilient Distributed Computing ​
Author: Parsa Moradi, Mohammad Ali Maddah-Ali
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2609.00552v1 Announce Type: new Abstract: Existing coded-computing designs do not explicitly exploit the intrinsic structure of the input data. In communication systems, statistical structure and redundancy are often removed through source coding (or compression) before channel coding is appli...
42. EEG-VID: Task-Guided Latent Predictive Pretraining for EEG Decoding and Assistive Target Selection ​
Author: Guanzhong Sun, Junyi Ma, Yuxuan Wu, Yanzi Miao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00566v2 Announce Type: new Abstract: We propose EEG-VID, a task-guided latent predictive pretraining framework for EEG decoding under session and subject shifts. EEG-VID predicts future latent EEG states from recent history using an exponential-moving-average target encoder and weak task ...
43. GeoPAR: Large-Scale Multi-Agent Combinatorial Optimization with Geometry-Guided Parallel Autoregressive Learning ​
Author: Wenjian Wu, Zesheng Jia, Jiaying Tang, Benyuan Yang, Jin Wang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2609.00577v1 Announce Type: new Abstract: Multi-agent combinatorial optimization problems are notoriously challenging due to their NP-hard nature. Recent parallel autoregressive neural solvers improve inference efficiency by allowing agents to make decisions simultaneously, but their performan...
44. CRAFT: Fine-Tuning Pre-hoc Explainability in AI-native 6G RAN ​
Author: Pranshav Gajjar, Vijay K Shah
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00590v1 Announce Type: new Abstract: The next generation of mobile networks is envisioned as fully AI-native, with AI-RAN architectures embedding small language models (SLMs) to perform reasoning over real-time telemetry. The state-of-the-art training paradigms for telecom LLMs, exemplifi...
45. Topological Steering ​
Author: Beno^it Gu'erand, Tan Minh Nguyen
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00597v1 Announce Type: new Abstract: With the rapid rise of large language models (LLMs), controlling undesirable model behaviors has become increasingly important. Existing behavioral control methods typically intervene directly in activation or feature space, but such approaches can be ...
46. Confess What You Know: Forget-Set Misalignment with Model Knowledge in LLM Unlearning ​
Author: Miso Kim, Georu Lee, Seungwon Jeong, Woojin Lee
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2609.00605v1 Announce Type: new Abstract: Machine unlearning for large language models (LLMs) often assumes that a pre-defined forget set matches what the model has memorized, but this frequently breaks in realistic privacy settings where the original training data is inaccessible. We term thi...
47. Breaking the Structural Identity: Personalized Federated LoRA Fine-tuning under Rank Heterogeneity ​
Author: Lei Wang, Jieming Bian, Letian Zhang, Jie Xu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00632v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable success across diverse domains, but their adaptation to privacy-sensitive, distributed datasets remains a challenge. While Federated Learning (FL) combined with Low-Rank Adaptation (LoRA) provides a...
48. DK-GBMKKM: Dynamic Kernel-Space Granular-Ball Multiple Kernel $k$-Means Clustering ​
Author: Xiaoyu Lian, Yuchao Zhang, Shuyin Xia, Siqi Zhong, Xuzhao Xiang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00647v1 Announce Type: new Abstract: Multiple kernel $k$-means integrates complementary nonlinear similarities by learning a combination of base kernels. Its pointwise optimization, however, is sensitive to noisy and boundary samples and repeatedly operates on sample-scale kernel matrices...
49. EEG-AS: Instance-Level Foundation Model Selection for EEG Foundation Models via Behavior Reconstruction ​
Author: Yunzhen Zhang, Ruoxi Piao, Hasan Onur Keles, Mustafa Misir
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00653v1 Announce Type: new Abstract: Electroencephalography (EEG) is a non-invasive technique for measuring neural activity and has been widely used in neuroscience applications. Recent advances in EEG foundation models have enabled strong performance across diverse neural decoding tasks....
50. HarmoCore: Functional Latent Diffusion for Sparse Reconstruction of Oscillatory Wave Fields ​
Author: Lihao Chen, Xinyu Zhang, Panqi Chen, Lei Cheng, Ting Zhang, Jianlong Li, Shikai Fang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2609.00679v1 Announce Type: new Abstract: Reconstructing oscillatory wave fields from scattered sensors is a severely underdetermined inverse problem. Beyond the challenges of general physical-field reconstruction, wave responses are complex-valued, frequency-sensitive, and highly oscillatory,...
51. A Study of Hidden-State Optimization Order in Predictive Coding Networks ​
Author: Xueyuan Li, Danilo Vasconcellos Vargas
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00686v1 Announce Type: new Abstract: Local learning methods offer an alternative to end-to-end backpropagation, but their unstructured local objectives can produce weak feature learning in deep networks. We study whether the order of hidden-state optimization can address this limitation. ...
52. Verdict Instability of OOD Scores under Reference Resampling ​
Author: Donghoon Lee, Shinjin Kang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2609.00691v1 Announce Type: new Abstract: Post-hoc out-of-distribution detectors are fitted on a finite reference set, so every score they produce is an estimate. If we had chosen a different set, some verdicts would have moved. We measure that movement by resampling the reference set and reco...
53. MUGEN: Generating Unlearnable Graph Examples for Multiple Learning Tasks ​
Author: Ziyan Liu, Chengshuai Zhao, Huan Liu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00696v1 Announce Type: new Abstract: Graph data across diverse domains can expose valuable relational information to unauthorized representation learning, creating a pressing need for protection against such misuse. Unlearnable examples offer a data-level defense by perturbing a training ...
54. Patterning in Practice: Debiasing Reward Models with Susceptibilities ​
Author: George Wang, Elizabeth Donoway, Daniel Murfet
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00699v1 Announce Type: new Abstract: Reward models trained on human preferences are known to suffer from length, formatting, and other stylistic biases. In this paper we use patterning, which reweights each preference pair according to its measured effect on posterior expectation values o...
55. Online Self-Weighted Fine-Tuning ​
Author: Haiquan Wen, Yiwei He, Bei Peng, Guangliang Cheng
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00734v1 Announce Type: new Abstract: Standard supervised fine-tuning (SFT) assigns the same explicit loss weight to every expert demonstration, regardless of the model's changing competence over training queries. Reinforcement learning (RL) based methods adapt update strength using model-...
56. Text Capability Loss in Vision-Language Adaptation: An Attention-Sink Diagnosis ​
Author: Minsik Choi, Geewook Kim, Young Geun Kim
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00746v1 Announce Type: new Abstract: Fine-tuning a pretrained LLM into a vision-language model (VLM) can erode the backbone's text capability, with the damage concentrated on tasks that require following exact output rules, such as instruction following, chain-of-thought reasoning graded ...
57. How Do Language Models Choose Between Context and Memory? ​
Author: Benjamin Shih, John Winnicki, Arianna Cao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2609.00753v1 Announce Type: new Abstract: When contextual information conflicts with the knowledge stored in model parameters, activation directions can be used to decode and steer which source the model follows. However, steering along a direction does not establish causality: whether the une...
58. Frozen Cores Need Task Signal: Fisher-Whitened Cross-Covariance for Low-Resource LLM Adaptation ​
Author: Wentao Ye, Zhanming Shen, Zhiqing Xiao, Yao Ding, Haobo Wang, Gang Chen
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00762v1 Announce Type: new Abstract: Parameter-efficient fine-tuning is usually framed as a question of how many parameters to update. Under a severe trainable-state budget, however, where those coefficients act is equally consequential. We study this choice through frozen-core adaptation...
59. Are You Thinking What I am Thinking? : Examining Conceptual Separation in Neural Architectures ​
Author: Jaee Ponde, Roshni Agarwal, Subhashis Banerjee
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00764v1 Announce Type: new Abstract: Neural networks are increasingly employed to identify both well-defined and ambiguous concepts, yet output-level metrics reveal little about how those concepts are represented internally. Our study asks if these networks exhibit \textit{conceptual sepa...
60. Subspace Levenberg Marquardt Algorithms in Training Neural Networks ​
Author: M. Duc Hoang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2609.00789v1 Announce Type: new Abstract: The Levenberg-Marquardt (LM) algorithm is a well-known second-order method for rapid convergence and strong robustness when training small- to medium-sized neural networks (NNs). However, its computational and memory costs increase significantly as the...
61. HarnessEvolve: Learning from Reference Trajectories for Reliable Agent Self-Evolution ​
Author: Wen Jiang, Mingmin Chu, Yimeng Tian, Qianxin Zhang, Haofei Yang, Rui Yang, Yang Liu, Tao Lv, Fangming Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.00829v1 Announce Type: new Abstract: Self-evolving agents advance toward autonomy by optimizing their harness---prompts, skills, tools, and execution logic---based on environmental feedback. This paradigm, however, is hampered by three challenges: \textit{credit assignment failure}, where...
62. Conditional Flow Matching for ML-Based Inverse Design Problems ​
Author: Juliana Felder, Milad Habibi, Soheyl Massoudi, Mark Fuge
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00863v1 Announce Type: new Abstract: Engineering inverse design is often limited by the high computational cost of iterative solvers for optimization problems constrained by partial differential equations (PDEs) and by their sensitivity to initialization. Deep generative models can produc...
63. MemoryWalker: Stop Training Agents on Contexts They Never Saw ​
Author: Zinco J, Xunjie Zhu, Shen Huang, Zhenyi Wang, Pengjun Xie, Jieping Ye
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2609.00865v1 Announce Type: new Abstract: Production agent harnesses such as Claude Code and Qwen-Agent compress context during rollout, but training under compression creates a conditioning problem: every eviction branches the effective history, so the learning object is a tree rather than a ...
64. iPINN for Broadband CARS Phase Retrieval: A Framework for Function Approximation and Inverse Modeling Problems in Nonlinear Spectroscopy ​
Author: Ravi Teja Vulchi, Carl Messerschmidt, Mohammadsadegh Vafaeinezhad, Rajendhar Junjuri, Tobias Meyer-Zedler, Juergen Popp, Thomas Bocklitz
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, physics.data-an
arXiv:2609.00883v1 Announce Type: new Abstract: Phase retrieval in broadband coherent anti-Stokes Raman spectroscopy (BCARS) is an ill-posed inverse problem. The Raman-like signal is encoded in the imaginary part of the resonant susceptibility, which mixes coherently with a non-resonant background (...
65. Poisson-Gamma Dynamical Systems with Time-varying Transition Dynamics ​
Author: Jiahao Wang, Yijun Wang, Nan Fang, Sikun Yang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.00896v1 Announce Type: new Abstract: Bayesian methodologies for handling count-valued time series have gained prominence due to their ability to infer interpretable latent structures and to estimate uncertainties. Among these Bayesian models, Poisson-Gamma Dynamical Systems (PGDSs) are pr...
66. When Metropolis and Hastings Meet Bradley and Terry: Exact MCMC From Preference Voting ​
Author: Ariel Smogorghevski, Nir Rosenfeld, Yaniv Romano
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, stat.CO, stat.ML
arXiv:2609.00905v1 Announce Type: new Abstract: Sampling from distributions conditioned on desired semantic properties is an emerging challenge in modern generative modeling. Metropolis-Hastings (MH) provides a principled route to conditional sampling, but requires access to exact pointwise target-d...
67. The Multiple Timescales of Gradient Descent on the Edge of Stability: A Perturbative Derivation of the Central Flow ​
Author: Rapha"el Berthier
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML
arXiv:2609.01034v1 Announce Type: new Abstract: The central flow of Cohen et al. (2025) is an empirically accurate continuous-time model of gradient descent at the edge of stability in deep learning, However, its derivation is heuristic. We propose a perturbative regime in which the central flow is ...
68. From Truncation to Commitment: Persistent Context in Uniform Discrete Diffusion ​
Author: Satoshi Hayakawa
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.PR, stat.ML
arXiv:2609.01043v1 Announce Type: new Abstract: Uniform-state discrete diffusion models update all tokens in parallel while keeping every position revisable. Even when the commonly used top-$p$ rule leaves only one candidate at a position, that choice affects only the current reverse step and can be...
69. SAGE: Subpopulation-Aware Generative Enhancement for Mitigating Spurious Correlations ​
Author: Yiming Luo, Rongqiang Zhao, Jie Liu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2609.01051v1 Announce Type: new Abstract: Spurious correlations pose a significant challenge to the robustness of modern machine learning. The inherent imbalance in dataset distributions often leads traditional Empirical Risk Minimization (ERM) models to rely on majority spurious attributes fo...
70. Let Confidence Change, Not the Prediction: Prediction-Preserving Repair for Post-hoc Calibration ​
Author: Daehwan Kim, Haejun Chung, Ikbeom Jang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2609.01072v2 Announce Type: new Abstract: Post-hoc calibration corrects reported confidence, yet a multiclass calibrator can also change the associated top-1 prediction. Accuracy captures only the net effect of these changes on correctness, not how often predictions change; the Top-1 Predictio...
71. Modelpedia: A Catalog of Model Findings for the Meta-Science of AI ​
Author: Franciszek Bernat (Centre for Credible AI, Warsaw University of Technology), Dawid P{\l}udowski (Centre for Credible AI, Warsaw University of Technology), Micha{\l} Jan W{\l}odarczyk (Centre for Credible AI, Warsaw University of Technology), Luca Longo (University College Cork), Jianlong Zhou (University of Technology Sydney), Andreas Holzinger (Human-Centered AI Lab), Riccardo Guidotti (University of Pisa, ISTI-CNR), Wojciech Samek (Technical University of Berlin, Berlin Institute for the Foundations of Learning and Data), Przemys{\l}aw Biecek (Centre for Credible AI, University of Warsaw)
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01090v1 Announce Type: new Abstract: Scientific knowledge about AI models is produced faster than the community can organize it. Every few months a new foundation model reshapes the field and hundreds of papers, blogs, and technical reports document how each behaves or fails. Yet, these f...
72. Subliminal Learning as Trait-Direction Drift: A Mechanism and Targeted Control under SFT Distillation ​
Author: Zhixuan Liu, Zhichen Dong, Yuyu Fan, Xiangtian Li, Chao Yang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01091v2 Announce Type: new Abstract: Beyond intended capabilities, model distillation can transfer hidden traits from a teacher. A teacher biased by a system prompt can generate semantically clean training data, such as numeric sequences, that still causes a downstream student to inherit ...
73. Neural Symbollic Regression Using Deep Learning and Sparse Modelling ​
Author: Ravi Kumar U, Sumitra S
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, cs.SC
arXiv:2609.01102v1 Announce Type: new Abstract: Symbolic Regression (SR) seeks to find succinct mathematical expressions that represent the fundamental relationships within data, providing interpretability and scientific understanding that exceeds that of black-box models. Nevertheless, traditional ...
74. Replicating TRACE: A Practitioner's Guide to Its Threshold and Particle Budget ​
Author: Alex Chadyuk, Alicia Zhang, Roy Kucukates
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01108v1 Announce Type: new Abstract: TRACE (Math & Lienhart, arXiv:2602.01135) reads causal graphs over event types out of a pretrained autoregressive sequence model by thresholding a per-position conditional-mutual-information estimate at a fixed tau. We independently replicate its headl...
75. When Does Online Adaptation Pay on the Edge? A Leakage-Free Evaluation of Warmup, Learning-Rate Selection, and Resource Trade-offs for Time-Series Forecasting ​
Author: Takumi Fujimoto, Hiroaki Nishi
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01126v1 Announce Type: new Abstract: Online adaptation can help edge time-series forecasting under distribution drift, but its measured benefit is sensitive to evaluation choices. We study six public multivariate streams, including building-sensor and smart-meter data, under a leakage-fre...
76. Scaled Idempotence in Transformer Attention: Paired OV Geometry and Shared-Value Algebras ​
Author: Jiming Feng, Junliang Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01129v1 Announce Type: new Abstract: We identify a recurrent algebraic regularity in Transformer attention: a sparse subset of effective OV operators $T=OV^\top$ nearly closes under composition, $T^2\approx\alpha T$. Across six pretrained endpoints spanning 2.8B--235B parameters, 3.98--8....
77. Superposed Latent Autoencoder ​
Author: Quanling Zhao, Jiaying Yang, Tianqi Zhang, Ziyang Hao, Fatemeh Asgarinejad, Flavio Ponzina, Tajana Rosing
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01158v1 Announce Type: new Abstract: Autoencoders typically meet tight latent-memory budgets by making each latent representation smaller, sacrificing representational capacity. We ask a different question: can multiple wider latents be stored together instead? We introduce the Superposed...
78. CopyShield: A Cross-Level Benchmark of Copyright Defenses in LLMs ​
Author: Maryam Alshehyari, Dushyant Singh Chauhan, Samuele Poppi, Martin Takac, Salem Lahlou, Nils Lukas
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01161v1 Announce Type: new Abstract: Large language models can reproduce memorized text verbatim, yet copyright defenses are usually evaluated under incompatible protocols. We introduce CopyShield, a controlled benchmark comparing three representative defenses at distinct intervention lev...
79. Pre-carved Niches: The Formation Dynamics of Modular Task Partitions in Early LLM Training ​
Author: Guangqi Li, Yongxin Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01170v1 Announce Type: new Abstract: Large language models exhibit a modular internal organization that mirrors well-studied functional networks of the human brain, but how this organization forms during training is unknown: prior work has characterized finished models, not the formation ...
80. Births are difficult to predict even with rich survey and full-population register data ​
Author: Elizaveta Sivak, Emily M. Cantrell, Thomas Emery, Javier Garcia-Bernardo, Flavio Hafner, Kasia Karpinska, Malte L"uken, Adrienne Mendrik, Joris Mulder, Hanzhang Ren, Varun Satish, Mark Verhagen, Angelica M. Maineri, Paulina Pankowska, Jasmin Abdel Ghany, Bruno Arpino, Giovanni Cassani, Julia Hellstrand, Katya Ivanova, Sanni Kuikka, Ana Macanovic, Charles Rahal, Felix C. Tropf, Roland J. Veen, Nicole Walasek, Dani"el van Wijk, Kelsey Q. Wright, Emilio Zagheni, Henry Abbink, Emanuele Aliverti, Matteo Amestoy, Tilbe Atav, Nicola Barban, Sunnee Billingsley, Goan J. Booij, Louis Boucherie, Yael Broos, Li Ya Chang, Jamie C. Chiu, Chiara Ludovica Comolli, Boris Cule, Qixiang Fang, Dennis M. Feehan, Rachel Ganly, Erwin Gielens, Rolando M. Gonzales Martinez, Andrea Gradassi, Rosember Guerra-Urzola, Mario Guerra-Urzola, St'ephane Guerrier, Enamul Hassan, Vincent A. Haverhoek, Andrew T. Hendrickson, Amber Howard, Yuxuan Jin, Sayash Kapoor, Erik-Jan van Kesteren, Iris ten Klooster, Marie Labussiere, Lydia T. Liu, Tiffany Liu, Adam Maghout, Simone Meneghello, Lasse Mohr, Clara H. Mulder, Saul J. Newman, Jessica Nis'en, Janis Norden, Mikkel Odgaard, Riccardo Omenti, Ozancan Ozdemir, Christina Pao, Paige Park, Gaia Penta, Juan C. Perdomo, Tanzir Pial, Alessio Piraccini, Federica Querin, Ziwei Rao, Christian Rellama, Adrien Remund, Frederieke Richert, Arnout van de Rijt, Mojtaba Rostami Kandroodi, Stijn J. Rotman, Lucas Sage, Germans Savcisens, Katrin Schwanitz, Steven Skiena, Alessandro Spata, Yannick Stadtfeld, Benedikt Stroebl, Gaetano Tedesco, Mathilde Theelen, Gianluca Tori, Abigail Tun-Mendicuti, Rishabh Tyagi, Keyon Vafa, Luiz Felipe Vecchietti, Linda Vecgaile, Willem R. J. Vermeulen, Maria-Pia Victoria Feser, Lionel A. Voirol, Thom B. Volker, Xinran Wang, Jiani Yan, Xinyi Zhao, Flora Zhou, Zuzana Zilincikova, Malvina Nissim, Matthew J. Salganik, Gert Stulp
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01194v1 Announce Type: new Abstract: Major life events have proven difficult to predict. Does this reflect limits of theory, data, and algorithms, or the large role of chance? We examine one outcome - having a child within three years - through a near-ideal setting for prediction: a data ...
81. Recent Developments in Transformer Inference Deployment on FPGA Platforms: A Survey ​
Author: Arjan Blankestijn, Uraz Odyurt, Amirreza Yousefzadeh
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AR
arXiv:2609.01212v1 Announce Type: new Abstract: With the rapid and continuous growth in the incorporation of machine learning models based on the Transformer architecture, capable deployment is in high demand. In this context, capable deployment refers to operational performance aspects, e.g., throu...
82. REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs ​
Author: Riyaaz Shaik, Chandru Venkataraman
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO
arXiv:2609.01215v1 Announce Type: new Abstract: Most vision-language-action (VLA) models -- OpenVLA, $\pi_0$, RT-2, RDT-1B -- are monolithic: they emit raw motor commands or short action chunks without organizing behavior into reusable abstractions, so they degrade on long-horizon tasks and resist i...
83. Multi-Head Self Attention is a Parameter Identification Mechanism ​
Author: W. Ross Morrow
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2609.01231v1 Announce Type: new Abstract: We prove that a multi-head scaled dot product attention can be viewed as a parameter identification strategy. The ratio of unidentified parameters to the total number of parameters scales like the reciprocal of the number of heads ($1/2 \to 1/(2H)$), m...
84. Post-Training Science for Supervised Fine-Tuning ​
Author: Charles O'Neill, Mudith Jayasekara, Harry Partridge
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2609.01244v1 Announce Type: new Abstract: Every supervised fine-tuning run forces the same chain of decisions, such as learning rate, batch size, LoRA or full fine-tuning, how many epochs, which optimiser, and what data to feed the model. Each of these is typically rediscovered from scratch fo...
85. Explore More, Drift Less: Outcome-Only Reinforcement Learning Can Suffice for Long-Horizon Interactive Agents ​
Author: Liming Pu, Xiaoxia Li, Yifu Liu, Teng Cao, Bin Yang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01245v1 Announce Type: new Abstract: Reinforcement learning is a natural way to post-train LLM agents for long-horizon interactive tasks judged only by end-of-task verification, yet a shared belief holds that outcome-only RL soon hits a ceiling on small open models. Recent work therefore ...
86. Solving In-Table Prediction Problems by Deep Neural Networks with Performance Evaluation Using Synthetic Data ​
Author: Xiao Zhao, Daniela Oelke
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01262v1 Announce Type: new Abstract: Tabular deep learning (TDL) leverages neural networks (NN) to extract patterns from tabular data. Traditional TDL methods follow a supervised learning paradigm, where a target feature is explicitly given. In this work, however, we explore a different a...
87. Position: Privacy Is a Claim, Not a Property of Synthetic Data ​
Author: Jiachen Zhao, Antonia Januszewicz, Taeho Jung
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2609.01273v1 Announce Type: new Abstract: Synthetic data has become a common component of machine learning research. While widely adopted, its use in privacy-sensitive contexts has quietly shifted from a claim of residual inference risk under stated assumptions to an appearance-based property ...
88. The Constitutional Coverage Trilemma in AI Governance ​
Author: Natalija Mitic, Soona Sedahmed A. O., Mamadou Selly Ly, Moustapha Cisse
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01275v1 Announce Type: new Abstract: Frontier AI systems function as \emph{constitutional institutions}: each deployed model encodes an implicit ranking among safety, helpfulness, honesty, autonomy, and equity. We ask whether the supply of frontier constitutional types covers human demand...
89. One-Layer Transformer Provably Learns Multiclass One-Nearest Neighbor in Context ​
Author: Skanda Athreya, Yutong Wang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2609.01311v1 Announce Type: new Abstract: We extend recent work establishing an equivalence between one-layer transformers and nearest-neighbor classifiers in the binary setting to the multiclass case. By leveraging the simplex encoding, we show that one-layer transformers with an argmax class...
90. Bandits in Prod: Hyperparameter Optimization at Inference Time ​
Author: Louis Abraham, Tuan-Anh Nguyen, Nicolas Devatine
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01335v2 Announce Type: new Abstract: Many production systems can assess a configuration only by using it on live requests and observing noisy feedback. Modern agentic systems are a prominent example, with inference-time choices such as model selection, retrieval depth, prompting strategy,...
91. SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers ​
Author: Shaowen Wang, Ge Zhang, Kairong Luo, Yuhao Wu, Shaofan Liu, Jiaheng Liu, Wenhao Huang, Shen Yan, Jian Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01343v1 Announce Type: new Abstract: Looped Transformers increase effective depth by iterating a shared block of layers, but most evaluations compare at fixed model size, conflating architectural advantage with extra FLOPs. We study looping on Mixture-of-Experts Transformers while closely...
92. Contribution-Aware Bandwidth Allocation for Multimodal Split Learning ​
Author: Iason Ofeidis, Leandros Tassiulas
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.NI
arXiv:2609.01406v1 Announce Type: new Abstract: Multimodal models are increasingly the default option for perception at the network edge, yet they are trained almost entirely in the datacenter, because a client holding several sensor streams cannot host an encoder per modality. Split Learning makes ...
93. Predicting Subsurface Abnormalities Growth using Physics-Informed Neural Networks ​
Author: Mehrdad Shafiei Dizaji, Hoda Azari
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01417v1 Announce Type: new Abstract: The research explores the pioneering integration of Physics-Informed Neural Networks (PINNs) into the domain of Ground-Penetrating Radar (GPR) data prediction. This research presents a detailed development framework for a specialized PINN model, profic...
94. Provably Safe Sim-to-Real Transfer ​
Author: Tingting Ni, Maryam Kamgarpour
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01418v1 Announce Type: new Abstract: To mitigate the sample complexity of real-world reinforcement learning (RL), a common practice is to first train a policy in a simulator, where samples are cheap, and then deploy the learned policy in the real world with the hope that it generalizes ef...
95. CATeye: Coupled Attribute-Topology Invariance Learning for Voucher Abuse Detection ​
Author: Tian Tian, Shuaicheng Niu, Hao Kuang, Yuanhang Hu, Dong Li, Zhiqi Shen
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01425v1 Announce Type: new Abstract: Voucher abuse poses a major challenge in e-commerce, where malicious users exploit promotional vouchers for profit. Unfortunately, fraud patterns evolve rapidly over time and across regions, causing distribution shifts that degrade existing detection m...
96. TRIAGE: Three-level Routing and Intelligent Agent Guidance for Efficient Execution ​
Author: Ruocan Wei
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01428v1 Announce Type: new Abstract: Large Language Model (LLM) agents based on the ReAct paradigm have demonstrated remarkable capabilities in tool use and task execution. However, ReAct suffers from a fundamental efficiency problem: every query triggers a complete reasoning loop from sc...
97. Learning Sparse Decision Trees via Transformer Variational Auto-Encoders ​
Author: Giacomo Fidone, Alessio Cascione, Riccardo Guidotti
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01430v1 Announce Type: new Abstract: Decision trees are among the most widely used models in machine learning, largely due to their transparent decision logic, making them well-suited for high-stakes decision-making contexts. However, most existing learning algorithms focus on predictive ...
98. Efficiently Estimating Optimal Hyperparameter Scaling Laws through Power-Law Entropy Search ​
Author: Zhiliang Chen, Sebastian Ament, David Eriksson, Maximilian Balandat, Eytan Bakshy, Jihao Andreas Lin
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01431v1 Announce Type: new Abstract: Optimal hyperparameter scaling laws describe how the best hyperparameters for large language model (LLM) training change with model and data scale, enabling practitioners to predict optimal configurations at production scales without expensive large-sc...
99. Edge-Girth as a Structural Edge Feature for Graph Neural Networks ​
Author: Lilian Marey, Charlotte Laclau
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01441v1 Announce Type: new Abstract: Graph neural networks (GNN) based on message passing are provably no more powerful than the one-dimensional Weisfeiler--Leman colour-refinement test (1-WL): two graphs it cannot tell apart receive identical representations, however deep or wide the net...
100. Diffusion as a Training Curriculum for Timestep-Free Iterative Reasoning ​
Author: Mariia Drozdova, Aidan Sirbu, Pietro Miotti, Robert Obryk, Mayalen Etcheverry, Eyvind Niklasson, Blake Richards
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01449v1 Announce Type: new Abstract: Diffusion models and recursive reasoners are both iterative, but they carry information across iterations differently. We add a persistent hidden state to a diffusion denoiser and remove its timestep conditioning, leaving a single shared update that ca...
101. Rethinking Learnability in Offline Data-driven Optimization ​
Author: Chao Qian, Chen-Guang Wang, Rong-Xi Tan, Ke Xue
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2609.01493v2 Announce Type: new Abstract: Black-Box Optimization (BBO) has broad applications, while traditional algorithms such as evolutionary algorithms and Bayesian optimization face efficiency challenges as real-world BBO problems grow increasingly complex. Data-driven optimization has be...
102. Optimizing Byzantine Node Placement in Decentralized Federated Learning ​
Author: Edoardo Gabrielli, Gabriele Tolomei
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01495v1 Announce Type: new Abstract: Security evaluations of decentralized federated learning (DFL) typically focus on how Byzantine participants behave, while largely overlooking which participants are compromised. Yet, because aggregation is distributed over a communication graph, the p...
103. LatentPress: Context Compression Beyond Text and Vision ​
Author: Zhengze Zhou, Hejian Sang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01507v1 Announce Type: new Abstract: Compressed context is usually carried as human-readable text or as rendered images that must be decoded, even when its consumer is a language model. We introduce LatentPress, which writes conversational histories and long documents into a third represe...
104. Quantum Sparse Autoencoders for Q-Matrix Estimation in Cognitive Diagnosis ​
Author: Arif Hassan Zidan, Yi Pan, Bowen Guo, Xiang Li, Yu Bao, Yingfeng Wang, Tianming Liu, Wei Zhang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01537v1 Announce Type: new Abstract: Q-matrices play a central role in cognitive diagnosis within educational data mining (EDM), specifying which latent skills each assessment item requires. Data-driven Q-matrix estimation remains challenging when assessments involve many correlated skill...
105. NashDreamer: Model-Based Reinforcement Learning for Zero-Sum Imperfect-Information Games ​
Author: Tom'a\v{s} Hole\v{c}ek, Viliam Lis'y
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01549v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) has achieved remarkable results in single-agent domains, yet its extension to competitive imperfect information games (IIGs) remains underexplored. In multi-agent settings, opponent-induced non-stationarity com...
106. A Mathematical Theory of Reusable Neural Bases for Network Compression ​
Author: Binshuai Wang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2609.01550v1 Announce Type: new Abstract: As large AI models become increasingly prevalent across a wide range of applications, memory cost has become a critical bottleneck in both training and inference. To mitigate this issue, we introduce the Linear Reusable Neural Bases Architecture (LRNBA...
107. Retrieved but not ranked: surface-form bias in structural retrieval, from mathematics to agent trajectories ​
Author: Nabira Rashid, Manolis Kellis
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR
arXiv:2609.01556v1 Announce Type: new Abstract: We evaluate embedding retrieval where surface form and meaning are pulled apart on purpose: retrieving items that share underlying structure but not wording, in two unrelated domains under one protocol, competition mathematics (MathNet-Retrieve; 500 qu...
108. Gradient-Update Mismatch: Rethinking Conflict-Free Training of Physics-Informed Neural Networks ​
Author: Jing Xiao, Xinhai Chen, Qinglin Wang, Menghan Jia, Zhiquan Lai, Dongsheng Li, Jie Liu, Tiejun Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2609.01558v1 Announce Type: new Abstract: Training Physics-Informed Neural Networks (PINNs) requires jointly optimizing physics residual and initial/boundary condition loss terms, which often induce conflicting gradients. Gradient surgery methods mitigate this issue by constructing directions ...
109. The Structure of Quantization Damage in LLMs: Why the Next Bit Should Be Spent Globally ​
Author: Jundong Hu, Shekar Ramachandran
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2609.01587v1 Announce Type: new Abstract: Post-training quantization (PTQ) is widely used to reduce the cost of serving large language models (LLMs), but its accuracy cost is uneven and is often tuned per model. We study where quantization damage occurs and how to allocate a small additional p...
110. I-CARE: Analysis of interference-related phenomena in a controllable, diverse and representative unlearning setting for text-to-image models ​
Author: Leonardo Santiago Benitez Pereira, Marcos Escudero Vi~nolo, Luis Herranz Arribas
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2609.00003v1 Announce Type: cross Abstract: Machine unlearning studies the removal of knowledge from an AI model, making the system forget a concept it previously learned. Despite rapid progress in generative machine unlearning, the unintended degradation of semantically related concepts that ...
111. ES-AHD: An Evolution Strategy Framework for Automatic Heuristic Design ​
Author: Yutao Lai, Kezhao Lai, Hai-Lin Liu, Yuping Wang, Ping Guo
Published: 9/2/2026, 4:00:00 AM
Categories: cs.NE, cs.LG
arXiv:2609.00023v1 Announce Type: cross Abstract: In this paper, we introduce ES-AHD, a novel framework that fundamentally integrates Evolution Strategy (ES) into Large Language Model (LLM)-driven Automatic Heuristic Design (AHD). Existing evolutionary approaches predominantly rely on random, indivi...
112. UI-Venus-2 Technical Report ​
Author: Venus Team, Zhuohan Cai, Haoxing Chen, Jiaxuan Chen, Weizhi Chen, Changlong Gao, Zhangxuan Gu, Yuan Guo, Yusong Hu, Jianrong Jiang, Jianguo Li, Runze Li, Jinzhen Lin, Zhenyu Ma, Changhua Meng, Han Peng, Xinyu Qiu, Shuheng Shen, Zhongyi Shui, Weiqiang Wang, Ming Wen, Zhuoer Xu, Hang Yan, Kaiwen Yang, Ruilin Yao, Nanjun Yu, Zhengwen Zeng, Lianrui Zhang, Yunzhu Zhang, Zhe Zhao, Beitong Zhou
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CV, cs.LG
arXiv:2609.00028v1 Announce Type: cross Abstract: Multimodal GUI agents have emerged as a promising paradigm for digital task automation, yet transitioning from benchmark-oriented models to dependable real-world applications remains challenging due to limited environment coverage, brittle task const...
113. Dense Weak Hiding: Closing Complexity Gaps in Nonconvex and PL Finite-Sum Optimization under Individual Smoothness ​
Author: Yuxing Peng, Zhiqing Tang, Weijia Jia
Published: 9/2/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, math.OC
arXiv:2609.00045v1 Announce Type: cross Abstract: Under individual smoothness, the optimal incremental first-order oracle (IFO) complexity of nonconvex finite-sum optimization has remained open. Known algorithms use $O(n+\sqrt{n},\Delta L_{\max}/\varepsilon^2)$ calls, while prior lower bounds miss ...
114. Towards Agentic Cloud Engineering: Graph and Loop Engineering with a Zero-Trust Agent Harness ​
Author: Sagar Srinivas Sakhinana, Venkataramana Runkana
Published: 9/2/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2609.00050v1 Announce Type: cross Abstract: Agentic AI is enabling cloud-based workflows in which autonomous agents reason over operational state, invoke authorized tools, modify software and infrastructure, deploy services, verify execution outcomes, and adapt across long-horizon, multistep t...
115. AgentProv: Auditing Agentic LLM API Providers via Tool-use Policy Probes ​
Author: Xun Wang, Bihe Zhao, Michael Backes, Franziska Boenisch, Adam Dziedzic
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG
arXiv:2609.00052v1 Announce Type: cross Abstract: Commercial LLM APIs advertise a specific foundation model, but the served backbone may be silently substituted, quantized, or wrapped, for example to save deployment costs. All existing audits decide backbone identity from the text-output channel, wh...
116. ValueGraph: Value-Signal Guided Graph Pre-training for Contextualized User Representation ​
Author: Yitong Han, Wei Gao, Yi Zhao, Prasanta Bhattacharya, Fengzhu Zeng, Mohammad Amanlou
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.00057v1 Announce Type: cross Abstract: Value signals are aggregated user-level moral representations that capture users' inferred value-related tendencies from their online discourse. User behavior on social media is shaped not only by what users say or whom they interact with, but also b...
117. OCGQuant: Outlier-Companion Grouping for NVFP4 Quantization ​
Author: Yishan Yao, Binjun Li, Hanling Yi, Pengyu Li, Xiaoqing Liu, Zihan Yang, Xiaotian Yu, Zhiwen Yu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.00066v1 Announce Type: cross Abstract: NVFP4 is an efficient microscaling format for low-bit inference, but activation outliers can still degrade quantization accuracy within NVFP4 blocks. Within each quantization block, large activations can dominate the block scale, increasing the quant...
118. When Prediction Error Is Not Enough: Evaluating Nuisance-Function Prediction for Causal Estimation ​
Author: Cong Cao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, stat.ME
arXiv:2609.00071v1 Announce Type: cross Abstract: Prediction error is widely used to evaluate nuisance-function estimators in causal inference, but its relationship with causal estimator performance may differ across performance measures. We studied this question in a partially linear model using Mo...
119. Different representation learning objectives recover distinct latent structures from the same psychometric data ​
Author: Cong Cao, Tassos C. Kyriakides, Pambos Vrasidas
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, stat.ME
arXiv:2609.00100v1 Announce Type: cross Abstract: Psychometric questionnaires contain rich item-level information, yet it remains unclear whether different representation learning objectives recover the same latent organization. We investigated this question using 757 matched teacher-child pairs fro...
120. Lingua Franca or Probing Artifact? Rethinking Latent Language in Multilingual LLMs ​
Author: Deniz Bayazit, Badr AlKhamissi, Antoine Bosselut
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.00155v1 Announce Type: cross Abstract: Latent language identification is often used to argue that multilingual language models route computation through language-specific states, such as English pivots. However, existing probes infer latent language from different signals, such as the geo...
121. Synthetic Worlds for Temporal Evaluation and Knowledge Updating in LLMs ​
Author: Jonathan Zheng, Zirui Shao, Alan Ritter, Wei Xu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2609.00184v1 Announce Type: cross Abstract: Large language models (LLMs) rely on static pretraining corpora, causing their knowledge to become outdated over time. Existing approaches for evaluating knowledge edits either suffer from rapid contamination or rely on counterfactual edits that conf...
122. Provably Efficient Federated Reinforcement Learning with Linear Function Approximation and Logarithmic Communication Cost ​
Author: Zihang Liang, Haochen Zhang, Lingzhou Xue
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2609.00193v1 Announce Type: cross Abstract: We study federated online reinforcement learning with linear function approximation. While recent multi-agent reinforcement learning algorithms achieve strong regret guarantees, they typically require sharing raw trajectories. This reliance incurs a ...
123. CompanionSim: Synthetic Data for Evaluating Anthropomorphism in Human-AI Relationships ​
Author: Jacy Reese Anthis, Mark D'iaz, Renee Shelby
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CL, cs.LG
arXiv:2609.00250v1 Announce Type: cross Abstract: Many people now see AI systems as not just productivity tools but as social companions. Researchers are eager to study the consequences of AI companionship behaviors, such as validation, which evoke trust, empathy, and attachment in human-human inter...
124. Exact Global MCMC with Denoising Diffusion ​
Author: Mitch Hill
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2609.00279v1 Announce Type: cross Abstract: This work shows that diffusion models learned with standard denoising loss can provide effective global MCMC proposals for complex high-dimensional target densities. The method is motivated by the observation that sequentially applying a forward and ...
125. Lightweight Adaptation of EEG Foundation Models for Stroke Motor Imagery Decoding: Domain Shift and Subject-Level Robustness ​
Author: Anh T. Nguyen, Zihua Sun, Michelle J. Johnson
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CE, cs.LG
arXiv:2609.00282v1 Announce Type: cross Abstract: Motor imagery (MI) electroencephalography (EEG) decoding could support post-stroke rehabilitation, but models developed on healthy cohorts may not transfer reliably to pathological EEG. We evaluated whether Low-Rank Adaptation (LoRA) can efficiently ...
126. WiSDoM: Wireless Sparse Decision Transformer with Mixture-of-Experts for Multi-Task Mobile Network Optimization ​
Author: Fatih Temiz, Shavbo Salehi, Melike Erol-Kantarci
Published: 9/2/2026, 4:00:00 AM
Categories: cs.NI, cs.AI, cs.LG
arXiv:2609.00284v1 Announce Type: cross Abstract: Emerging 6G wireless networks are expected to operate across diverse deployment scenarios, where variations in network topology, user mobility, traffic demand, and radio conditions challenge the scalability of conventional radio resource management (...
127. TRUST: Threshold-Recalibrated Uncertainty-Safe Training for Certified Dismissal in Breast Cancer Screening ​
Author: Parham Hajishafiezahramini, Matthew Hamilton, Edward Kendall, Gregory Doyle, Oscar Meruvia Pastor
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2609.00300v1 Announce Type: cross Abstract: Reducing the review of clearly cancer-negative screening mammograms could lower radiologist workload without compromising cancer detection. We propose a closed-loop threshold-aware training strategy in which the dismissal threshold is recalculated du...
128. Workload Identification with Physical Side Channels for AI Governance ​
Author: Simone Gargiulo, Gabriel Kulp
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CY, cs.LG
arXiv:2609.00309v1 Announce Type: cross Abstract: AI compute verification is one of the first tangible and tractable points for international policy aimed at AI governance. Determining whether frontier labs, or any operator, comply with agreements requires the regulating authority to discern how the...
129. The Curse of Multilinguality in Lexical Normalization ​
Author: Saman Rahbar
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.00329v1 Announce Type: cross Abstract: Lexical normalization rewrites the noisy, non-standard words that fill user-generated text (tmrw, u, gr8) into their standard forms. Because labelled data is scarce for most languages, a popular shortcut is to train a single model on many languages a...
130. Topic Matching in the Wild: Benchmark and Lessons from Real-World ASR Transcripts ​
Author: Saman Rahbar, Xiliang Zhu, Irvin Cardoza, David Rossouw
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.00330v1 Announce Type: cross Abstract: In contact centers, real-time agent-assist tools determine, for each of many predefined topics, whether a live customer utterance is relevant and display a coaching card to the agent when it is. The input is noisy and challenging: ASR(Automatic Speec...
131. Latent-Space No-Arbitrage Geometry of Generative Models for Implied Volatility Surfaces ​
Author: Jing Wang, Shuaiqiang Liu, Cornelis Vuik
Published: 9/2/2026, 4:00:00 AM
Categories: q-fin.CP, cs.AI, cs.LG, cs.NA, math.NA
arXiv:2609.00332v1 Announce Type: cross Abstract: Generative models for implied volatility surfaces must produce outputs that satisfy static no-arbitrage constraints. We study these constraints in latent space. For a fixed generator, we assign each latent code a scalar margin determined by the no-ar...
132. A Stable Aggregation Method for Quantum Federated Learning ​
Author: Shanika Nanayakkara, Shiva Raj Pokhrel
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2609.00356v1 Announce Type: cross Abstract: Quantum federated learning (QFL) enables clients to train quantum neural network (QNN) models without sharing private data. We find that aggregation in QFL is unstable under heterogeneous data, unreliable communication, variable fidelity, latency, an...
133. Dr. Claw: An AI Scientist Workspace for Vibe Research ​
Author: Dingjie Song, Hanrong Zhang, Dawei Liu, Yixin Liu, Zongxia Li, Zhengqing Yuan, Siqi Zhang, Henry Peng Zou, Zhiling Yan, Yuxuan Zhang, Yanfang Ye, Philip S. Yu, Lichao Sun
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CV, cs.LG
arXiv:2609.00365v1 Announce Type: cross Abstract: Command-line coding agents (e.g., Claude Code, Gemini CLI) can already read and write files and sustain long sessions, yet end-to-end research still fragments across chat tools, IDEs, terminals, and writing environments, and the decisions that make i...
134. Neurosymbolics for Data Engineering: Achieving Long Context Token Reduction Without Finetuning ​
Author: Vishvesh Bhat
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.00367v1 Announce Type: cross Abstract: Large Language Models are increasingly deployed for sophisticated data engineering tasks such as generating structured queries from natural language, Text-to-SQL, and automating complex spreadsheet operations. However, maximizing their utility demand...
135. Towards unsupervised representation learning for quantum data: quantum models with inference and generation ​
Author: Robin Lorenz, Eric Brunner, Marcello Benedetti
Published: 9/2/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2609.00372v1 Announce Type: cross Abstract: With quantum sensors, simulators and networks emerging, a future of quantum technology may produce quantum states as data---that is, coherently rather than as classical measurement records---thus motivating the study of suitable quantum generalisatio...
136. Risk-Aware Decision-Making for Autonomous Overtaking: A World Model-Based Mixture-of-Experts Framework ​
Author: Yongzhi Liu, Sunan Zhang, Jinchang Xu, Jiawei Wang, Yushu Qiu, Chen Lv, Weichao Zhuang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2609.00385v1 Announce Type: cross Abstract: Autonomous highway overtaking demands foresighted decision-making to handle complex interactions, stochastic traffic evolution, and temporal risk accumulation. However, standard safe reinforcement learning approaches typically rely on implicit value-...
137. Hidden relationships in a document-derived property graph: top-k chunk embeddings and inverse-distance weighting over a dynamically evolving ontology ​
Author: Bilge Kaan Karamete, Hunter Casten
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CE, cs.LG
arXiv:2609.00387v1 Announce Type: cross Abstract: Large language models extracting knowledge graphs from text capture only explicitly stated facts, often leaving semantically related entities disconnected across documents. We present an additive, engine-neutral second pass that discovers these laten...
138. NeuroPriv: Adversarial Representation Learning for Privacy in Wearable EEG Systems ​
Author: Sarmistha Sarna Gomasta, Bhawana Chhaglani, Prashant Shenoy
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CR, cs.HC, cs.LG
arXiv:2609.00390v1 Announce Type: cross Abstract: Wearable EEG systems may expose sensitive information beyond their intended health function, creating substantial risks to neuroprivacy. In this work, we show that commonly used EEG features can reveal participant identity and demographic attributes ...
139. A convolutional framework for detecting event-driven dynamics in energy price series ​
Author: Caixia Xu, Piotr Fryzlewicz
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2609.00402v1 Announce Type: cross Abstract: This paper develops a general convolutional neural network (CNN) framework for detecting heterogeneous event-driven dynamics in univariate time series windows. We show that the induced CNN class exactly represents classifiers based on range, maximum ...
140. DynaNDE: Dynamic Near-Data Expert Scheduling for Batched MoE Inference ​
Author: Xiaoyang Lu, Belthangady Akash Vi Narayana Pai, Xian-He Sun
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2609.00407v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models enable efficient scaling of large language model (LLM) inference but suffer from substantial data-movement overhead when deployed on neural processing unit (NPU)-based systems. Near-Data Processing (NDP) provides a pro...
141. Accelerating Chemical Kinetics for Exoplanet Atmospheres using Neural Networks ​
Author: Isaac Malsky, Xi Zhang, Tiffany Kataria, Matthew Graham, Ziyu Huang, Boris Bonev, Shang-Min Tsai, Elspeth K. H. Lee
Published: 9/2/2026, 4:00:00 AM
Categories: astro-ph.EP, cs.LG
arXiv:2609.00428v1 Announce Type: cross Abstract: Observations increasingly reveal the coupled radiative, chemical, and dynamical processes that shape exoplanet atmospheres. Interpreting these atmospheres requires models that can capture this complexity. However, multidimensional models remain funda...
142. SAGE: State-Grounded, Abstention-Aware Evaluation of Task-Oriented Dialogue Agents ​
Author: Rayan Khoury, Shih-Yao Lin, Pratyush Mishra
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2609.00434v1 Announce Type: cross Abstract: Evaluating task-oriented dialogue agents requires judging not merely whether a reply reads well but whether each turn advances the underlying workflow state correctly--a distinction conventional holistic LLM judges can miss because they evaluate the ...
143. Physiological Information Reliability: Cross-Layer Adaptive Resource Allocation for Cardiovascular Sensing ​
Author: Navaneeth Krishnan Kamalakannan, Janakiraman Kamalakannan, Harinisri Velmurugan
Published: 9/2/2026, 4:00:00 AM
Categories: eess.SP, cs.LG, cs.NI
arXiv:2609.00435v1 Announce Type: cross Abstract: Cardiovascular sensing systems must preserve clinically useful information despite signal degradation, wireless losses, energy constraints, and edge-computation latency. We introduce Physiological Information Reliability (PIR), a cross-layer framewor...
144. Capability-Gated Language Models: Security Composes, Utility Does Not ​
Author: Patrikas Vanagas, Augustas Ma\v{c}ijauskas, Laurynas Lopata
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2609.00445v1 Announce Type: cross Abstract: Deployed language model safeguards (safety fine-tuning, filtering, unlearning) vary by principal only outside the model weights: filters are reconfigured, tiers are multiplied, and artefacts are reissued; inside one set of weights every request meets...
145. Fractal dimension predicts quantum kernel collapse in angle-encoded data ​
Author: Ana Paula Appel
Published: 9/2/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2609.00475v1 Announce Type: cross Abstract: Angle-encoded quantum kernels on tabular data collapse when the feature map is wider than the intrinsic dimension of the data. We propose the correlation fractal dimension D2 as an a priori qubit budget: encode D2 coordinates chosen by FD-ASE instead...
146. Are Near-Tied LLM Rankings Robust to Family-DIF-Guided Benchmark Recomposition? ​
Author: Qiaoyuan Zheng, Yiqu Yang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2609.00482v1 Announce Type: cross Abstract: Small leaderboard gaps are often interpreted as evidence that one language model is better than another, but their sign may depend on which benchmark items are included. We test this using item-level responses from five benchmarks and a family-label-...
147. EvoFlint: An Evolutionary Atlas of Multi-Turn LLM Vulnerabilities ​
Author: Feitong Qiao, Liren Peng, Shiming Ren, Aishwarya Jadhav, Arghavan Bahadorinejad, Marinette Chen, Muhan Zhang, Abdulaziz Suria, Gennevi Lu, Anish Das Sarma
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CR, cs.LG
arXiv:2609.00487v1 Announce Type: cross Abstract: Frontier language models that refuse harmful single-turn prompts often comply when the same intent is reached gradually over many turns, making multi-turn attacks one of the least understood failure modes of large language models. Most automated red-...
148. Independent Reinforcement Learning in Discounted Markov Games ​
Author: Asrin Efe Yorulmaz, Ugur Aydin, Tamer Basar
Published: 9/2/2026, 4:00:00 AM
Categories: cs.GT, cs.AI, cs.LG, cs.SY, eess.SY, math.OC
arXiv:2609.00504v1 Announce Type: cross Abstract: In this work, we study radically uncoupled learning in discounted general-sum Markov games. Assuming ``$\mathsf{ETH}$ for $\mathsf{PPAD}$", we show that, for every fixed discount factor, there is no polynomial-time algorithm for computing inverse-pol...
149. Soft-Argmax for the Projective Plane via the Veronese Embedding ​
Author: Benjamin El-Zein, Dominik Eckert, Paul Zech, Christopher Syben, Bernhard Geiger, Steffen Kappler, Sebastian Stober
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2609.00521v1 Announce Type: cross Abstract: From horizon detection to fibre structures in X-ray imaging, many vision tasks recover lines via peak detection in Hough space $H=S^1\times\mathbb{R}$, the domain of orientation-offset pairs $(\theta,\rho)$. Differentiable pipelines extract coordinat...
150. EM^2Mem: Event-Centric Multimodal Memory for Large Language Models ​
Author: Yijun Chen, Yaqi Zheng, Yanya Li, Boyi Xiao, Buqiang Xu, Shuofei Qiao, Jizhan Fang, Xinle Deng, Yunzhi Yao, Xuehai Wang, Liuxin Zhang, Hui Li, Huajun Chen, Shumin Deng
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.MM
arXiv:2609.00551v1 Announce Type: cross Abstract: Multimodal memory offers a scalable interface for long-video question answering, but existing methods often retrieve captions, frames, transcripts, summaries, or graph facts as isolated fragments. Although searchable, such fragments are not generatio...
151. VoiceLongMemEval: Do Assistants Remember How You Sounded? ​
Author: Ramit Pahwa, Parivesh Priye, Apoorva Beedu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2609.00570v2 Announce Type: cross Abstract: With the growing scale of multi-agent architectures and large language models, deployed AI assistants are increasingly tasked with reasoning over long, continuous, multi-session conversation histories. Current benchmarks evaluate this dialogue histor...
152. Real-Time Neuromorphic Spectrum Intelligence Simulator ​
Author: Navaneetha Krishnan Kamalakannan
Published: 9/2/2026, 4:00:00 AM
Categories: eess.SP, cs.LG, cs.NI
arXiv:2609.00585v1 Announce Type: cross Abstract: We present the Real-Time Neuromorphic Spectrum Intelligence Simulator (RT-NuSIS), a modular framework to study spiking neural network (SNN) and memristor-inspired agents for dynamic spectrum access under constrained energy budgets and adversarial con...
153. BeamRMX: Radiation-Pattern-Driven Learning for Generalizable Beam Radio Map Prediction and Beam Management ​
Author: Yue Zhang, Xiucheng Wang, Wenshuo Chen, Nan Cheng
Published: 9/2/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2609.00615v1 Announce Type: cross Abstract: The evolution toward sixth-generation (6G) wireless networks is driving larger antenna arrays and highly directional multi-beam transmission, making accurate knowledge of beam-dependent spatial coverage important for beam management and environment-a...
154. Disciplined Bilevel Programming ​
Author: Hao Zhu, Joschka Boedecker
Published: 9/2/2026, 4:00:00 AM
Categories: math.OC, cs.CE, cs.LG, cs.MS
arXiv:2609.00644v1 Announce Type: cross Abstract: Bilevel optimization provides a natural modeling language for hierarchical decision problems. However, applying existing numerical solvers usually requires substantial manual analysis and reformulation. In this paper, we introduce disciplined bilevel...
155. Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search ​
Author: Enrong Pan, Ryan Zhou, Ting Hu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NE
arXiv:2609.00652v1 Announce Type: cross Abstract: Language model agents increasingly propose actions, observe external feedback, and explain their own behavior. Their confidence and rationales are convenient monitoring signals, but convenience is not verification. We introduce an environment-grounde...
156. Controllable Image Captioning with Prompt-Conditioned Scene Rewards ​
Author: Jongyeop Hyun, Taeyoung Kim, Hyounghun Kim
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG
arXiv:2609.00709v1 Announce Type: cross Abstract: Large Vision-Language Models produce fluent image descriptions but offer limited semantic control: users cannot reliably specify whether captions should emphasize attributes, relations, or particular image regions. We present Fine-grained Captioning ...
157. Prediction-Assisted Pricing and Admission for LLM APIs with Stochastic Token Consumption ​
Author: Patrick Wong
Published: 9/2/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2609.00710v1 Announce Type: cross Abstract: An LLM application often sells or internally allocates several service products: a small or premium model, a short or long token cap, and possibly multiple posted prices. The operational decision is not merely which model answers a prompt. A price ch...
158. MaskCode: Mask Transformer for Feedback-Assisted Coding With Linear Block Codes ​
Author: Jonggyu Jang, Hongjae Nam, Vishrant Tripathi, David J. Love, Hyun Jong Yang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, eess.SP, math.IT
arXiv:2609.00715v1 Announce Type: cross Abstract: Feedback-based coding schemes have demonstrated substantial performance gains over today's open-loop coding schemes. Unfortunately, these gains are usually achieved in idealized settings with perfect feedback. Over the last few years, machine learnin...
159. SOVER: Formal Certification of Optimization Reformulations via LLM-Assisted SMT Verification ​
Author: Swapnil Bhattacharyya, Mayank Baranwal
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, math.OC
arXiv:2609.00728v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown remarkable promise in translating and reformulating complex mathematical optimization problems across modeling languages. However, validating such transformations through empirical solver executions alone is un...
160. Agentic Empirical Asset Pricing: Methodological Foundations ​
Author: Yingjian Pan, Xiaowei Ding, Kay Giesecke
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-fin.ST
arXiv:2609.00731v1 Announce Type: cross Abstract: Recent advances in LLM agents enable a new paradigm for asset pricing, which we call Agentic Empirical Asset Pricing (AEAP): systems that autonomously conduct the scientific discovery process itself. We define AEAP and identify its core building bloc...
161. Semi-Supervised Classification with Informative Missing Labels in Weibull Mixture Models ​
Author: Jinran Wu, You-Gan Wang, Geoffrey J. McLachlan
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2609.00774v1 Announce Type: cross Abstract: We consider semi-supervised classification from a partially classified sample arising from a two-component Weibull mixture. The feature is observed for all data, whereas some class labels are missing. The probability of a missing label is modelled as...
162. Dense Process Supervision for Search Agents via Fact Utility Estimation ​
Author: Rongzhi Zhu, Xiangyu Liu, Yi Liu, Shuo Zhang, Ruirui Zhang, Rui Wu, Tao Jiang, Zequn Sun, Wenhao Xu, Wei Hu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2609.00833v1 Announce Type: cross Abstract: Reinforcement learning (RL) for search agents typically relies on outcome rewards. However, it often fails to achieve effective credit assignment, due to the unclear value of intermediate steps. It is hard to separate their contributions from the fin...
163. A Checklist to assess the energy and carbon impacts of ML/AI applications in Earth System Modeling ​
Author: Filippo Dainelli, Amirpasha Mozaffari, Marina Casta~no, Aina Gaya i `Avila, Llu'is Palma Garcia, Alessio Melli, Oscar Dimdore Miles, Amanda Duarte
Published: 9/2/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.AI, cs.LG
arXiv:2609.00847v1 Announce Type: cross Abstract: As machine learning and artificial intelligence find their way into nearly every aspect of climate, weather, and Earth system modeling, it is worth pausing to consider what our design decisions imply for the science and for the computational resource...
164. Does Fault Localization Beat a Fresh Attempt? A Placebo-Controlled Study of Test-Guided Code Repair ​
Author: Anik Jha
Published: 9/2/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2609.00854v1 Announce Type: cross Abstract: Fault localization can focus a code model's repair on the statements a failing test implicates, but a targeted edit may succeed merely because it is small, and a second model call may succeed without using the failure at all. We separate these explan...
165. The Visual Insensitivity Gap: Diagnosing When Vision-Language Models Fail to Use Visual Evidence ​
Author: Genpei Zhang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG
arXiv:2609.00868v1 Announce Type: cross Abstract: Vision-language models are evaluated by aggregate accuracy on multimodal benchmarks, a practice that implicitly assumes the model uses its visual input. We show this assumption fails on 40%--97% of samples across six VLMs and three perceptual benchma...
166. Sharp Mixed Spectral Barron Regularity of Coulombic Many-Electron Wave Functions ​
Author: Pingbing Ming, Hao Yu
Published: 9/2/2026, 4:00:00 AM
Categories: math.AP, cs.LG, cs.NA, math.NA
arXiv:2609.00872v1 Announce Type: cross Abstract: We establish sharp mixed spectral Barron regularity for eigenfunctions of molecular Coulomb Hamiltonians. The mixed norm is a Fourier $L^1$ norm with one isotropic weight and coordinate-product weights, and therefore detects regularity invisible to t...
167. FractalNet-Based Heterogeneous Federated Learning for Orbital Edge Intelligence in Satellite Mega-Constellations: A Wildfire Case Study ​
Author: Sai Puppala, Koushik Sinha
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.DC, cs.ET, cs.LG
arXiv:2609.00875v1 Announce Type: cross Abstract: Satellite mega-constellations are emerging as large-scale sensing, communication, and computation fabrics, yet their learning architectures remain largely inherited from terrestrial federated learning and ground-centric mission operations--- ill-suit...
168. Denoising Diffusion Generative Models Secretly Calculate Attentions ​
Author: Farzan Haddadi, Leila Monfared, Ebrahim Rezaii, Mohammadreza Malek-Mohammadi, Pejman Zakalvand, Narges Mokhtari
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG, cs.NE
arXiv:2609.00885v1 Announce Type: cross Abstract: Denoising diffusion models are the dominant architecture for image generation, whereas most natural language generation and modeling are primarily handled by well-known transformer architectures employing attention mechanism. Here, we show that diffu...
169. Vision-Language-Guided Pseudo-Labels for Unsupervised Domain Adaptation in Semantic Segmentation for Waste Sorting ​
Author: Udo Schlegel, Shubhangi, Gabriel Dax, Sai Rahul Kaminwar, Florian Karl, Thomas Seidl
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2609.00898v1 Announce Type: cross Abstract: Obtaining labeled data for semantic segmentation in applied settings (e.g., autonomous driving, industrial waste sorting) is expensive and often infeasible at scale. We present a cross-modal pseudo-labeling pipeline that enables unsupervised domain a...
170. Direct Optimization of a 3D Finite-Source Reflector via Neural-Network Parameterization ​
Author: Roel Hacking, Lisa Kusch, Martijn Anthonissen, Wilbert IJzerman
Published: 9/2/2026, 4:00:00 AM
Categories: physics.optics, cs.LG
arXiv:2609.00899v1 Announce Type: cross Abstract: We present a direct optimization method for three-dimensional freeform reflectors that transform the light of a finite-'etendue source into a prescribed far-field angular intensity distribution. The reflector profile is represented by a small neural...
171. Context-Grounding Gains Are Mediated by Pre-existing Machinery: Auditing GRPO, SFT, and DPO ​
Author: Prakhar Gupta, Vaibhav Gupta
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.00925v1 Announce Type: cross Abstract: Language models can ignore prompt evidence when it conflicts with memorized knowledge. Post-training can make models follow such evidence more reliably, but it is unclear whether these gains require new machinery or strengthen machinery already prese...
172. DualStake: Dual-Path Confidence Calibration in Deep Research Agents ​
Author: Yinuo Xu, Yuwei Liang, Jianjie Cheng, Meng Wang, Yongcan Yu, Shuo Lu, Jian Liang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.00935v1 Announce Type: cross Abstract: Deep Research agents tackle knowledge-intensive tasks through multi-round retrieval and decision-oriented generation. However, these agents suffer from severe overconfidence, making their expressed confidence unreliable for user trust and downstream ...
173. Embedded Conditional Independence Tests for Large Language Model Generated Text with an Application to German Parliament Speeches ​
Author: Marco Simnacher, Georg Keilbar, Benjamin K"onig, Christoph Lippert, Sonja Greven
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, math.ST, stat.ME, stat.TH
arXiv:2609.00946v1 Announce Type: cross Abstract: Conditional independence tests (CITs) test for conditional dependence between two random objects $X$ and $Y$ given a third random object $Z$. Existing CITs have limited applicability to high-dimensional data, especially multimodal data like text. How...
174. Right Frame, Wrong Rule: Cultural Cues Expose the Financial Knowledge Gap They Were Meant to Close ​
Author: Rania Elbadry, Ahmed Heakl, Saeed Almheiri, Fan Zhang, Muhra AlMahri, Xueqing Peng, Mohsinul Kabir, Shuyao Wang, Yi Han, Saadeldine Eletter, Duzhen Zhang, Preslav Nakov, Yuxia Wang, Fajri Koto, Zhuohan Xie
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.00999v1 Announce Type: cross Abstract: When a question has valid answers under different normative frameworks, a language model must decide which framework to use and whether it can answer correctly within it. We call this setting normative pluralism and study it in Islamic finance using ...
175. SinkPruner: Sink-Free Visual Token Pruning for Multimodal Large Language Models ​
Author: Shiyu Li, Zi-Yuan Hu, Shijia Huang, Yanyang Li, Yiwu Zhong, Liwei Wang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG
arXiv:2609.01004v1 Announce Type: cross Abstract: Despite their strong multimodal understanding ability, multimodal large language models (MLLMs) incur substantial computational overhead when processing long visual token sequences. To reduce inference costs, recent studies have explored visual token...
176. Web Price Extraction: State of the Art and an Adaptive Browserless Implementation ​
Author: Evgeniia Kositsyna, Jorge Lloret-Gazo
Published: 9/2/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, cs.NE
arXiv:2609.01030v1 Announce Type: cross Abstract: Price extraction from websites is a key task for market monitoring, price comparison, and business analytics in e-commerce. Existing approaches can be broadly divided into four groups, and understanding their trade-offs in accuracy and scalability is...
177. Spawn Freely, Act Sparingly: Progressive Risk Vesting for Recursive LLM-Agent Trees ​
Author: Molly Wang (Imperial Business School)
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, math.PR
arXiv:2609.01035v1 Announce Type: cross Abstract: Recursive LLM agents can broaden their search by spawning specialists. Some branches later request tools that send data or deploy code. When should a branch receive authority to act? We distinguish sandbox spawning, in which external controls prevent...
178. ViTAMINS: An Empirical Study of Training Self-Supervised Vision Transformers with Synthetic Hard Negatives ​
Author: Nikos Giakoumoglou, Andreas Floros, Kleanthis-Marios Papadopoulos, Tania Stathaki
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2609.01041v1 Announce Type: cross Abstract: We introduce ViTAMINS, a method that integrates synthetic hard negatives into unsupervised vision transformer pretraining to improve representation quality. Our approach is thoroughly benchmarked on ImageNet and transfer learning, image retrieval, co...
179. Accelerating Reinforcement Learning via MPC Solver-Gradient Guidance for Weights-varying MPC ​
Author: Baha Zarrouki, Arslan Thobani, Jasper Hoffmann, Mattia Piccinini, Rudolf Reiter, Felix Jahncke, S'ebastien Gros, Davide Scaramuzza, Johannes Betz
Published: 9/2/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SY, eess.SY
arXiv:2609.01061v1 Announce Type: cross Abstract: In Model Predictive Control (MPC), cost-function weights shape closed-loop behavior, yet changing conditions often make fixed parametrizations suboptimal and motivate context-dependent online adaptation. Learning such policies is difficult because be...
180. Artificial Rosetta Stone: Constrained Maximum A Posteriori (MAP) Reconstruction of Symbolic Raga Sequences via Order-k Markov Models ​
Author: Saanvi Raghavendran (Abstract Math Institute), Abhishek Bhattacharjee (Abstract Math Institute)
Published: 9/2/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, stat.ML
arXiv:2609.01064v1 Announce Type: cross Abstract: Reconstructing a damaged musical fragment is an inverse problem: the observed sequence contains partial information, while a raga encodes constraints limiting allowable completions. This paper formalizes a mathematical framework for this, proposing t...
181. From Language to Behavior: Scaling Sequence Transformers for Industrial Recommendation Ranking with Rec-Native Designs ​
Author: Jie Chen, Xiangqian Yu, Yanchao Lian, Tan Lu, Run Yang, Zhengchun Shang, Xing Wang, Cheng Chen, Ke Hu, Qiang Li, Tianjiu Yin, Xiaobing Liu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG
arXiv:2609.01240v1 Announce Type: cross Abstract: Scaling Transformers has driven large gains in language modeling, but transplanting this to behavior-sequence modeling in production ranking is challenging: recommendation differs in signal quality, where behavior sequences are noisy, temporally irre...
182. Relational Task Generation Language: A Declarative Specification Framework for Relational Deep Learning ​
Author: Oleksii Kolesnichenko, Jakub Pele\v{s}ka, Gustav \v{S}'{\i}r
Published: 9/2/2026, 4:00:00 AM
Categories: cs.PL, cs.DB, cs.LG
arXiv:2609.01292v1 Announce Type: cross Abstract: Relational Deep Learning (RDL) has become a powerful paradigm for learning from multi-tabular data. However, manually defining RDL prediction tasks is a laborious process that frequently results in data leakage. To address this issue, we introduce Re...
183. GazeRefine: Expert Gaze as a Test-Time Prompt for Training-Free Medical Image Segmentation ​
Author: Mohammed Oussama Benyahia, Marouane Tliba, Mohamed Amine Kerkouri, Taifour Yousra, Bin Wang, Max Bengtsson, Gorkem Durak, Elif Keles, Zuheng Ming, Marek Penhaker, Azeddine Beghdadi, Ulas Bagci, Aladine Chetouani
Published: 9/2/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.HC, cs.LG
arXiv:2609.01310v1 Announce Type: cross Abstract: Medical image segmentation remains difficult to scale because high-performing methods typically rely on dense expert annotations and task-specific training. We introduce GazeRefine, a training-free framework that uses gaze as an inference-time prompt...
184. MIDR: Enrichment-Augmented Indexing for Multimodal Document Retrieval ​
Author: Debanjan Mahata, Atharva Tendle, Daniel Preotiuc-Pietro, Yong Zhuang, Ozan Irsoy
Published: 9/2/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.CV, cs.LG
arXiv:2609.01316v1 Announce Type: cross Abstract: Retrieval over visually rich documents has a representation problem: important content often lives in tables, charts, figures, and layout relations that plain OCR linearizes, corrupts, or omits. ColPali-family visual retrievers address this with patc...
185. Matched Queries for Curvature and Density at Branching Junctions ​
Author: Ziqi Zhao, Qingjian Ni
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2609.01319v1 Announce Type: cross Abstract: At a junction, a score field can reveal weighted tangent rays, yet these first-order quantities do not determine how individual branches bend or how their densities change away from the center. Recovering this missing information is necessary for des...
186. Exploring Sparse Autoencoders in Text-Based Causal Confounding Adjustment ​
Author: Mian Zhong, Katherine A. Keith, Anjalie Field
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2609.01322v1 Announce Type: cross Abstract: In many settings, studying causal questions based on text data requires adjusting for confounding information within texts. Yet there is a tradeoff in constructing text representations for adjustment: they must be sufficiently large and/or dense to p...
187. mzCache: On-Device LLM Memory Management under Multitasking ​
Author: Hongseung Yu, Minsung Kim, Jongseok Park, Kyunghan Lee
Published: 9/2/2026, 4:00:00 AM
Categories: cs.OS, cs.DC, cs.LG
arXiv:2609.01338v1 Announce Type: cross Abstract: On-device mobile Large Language Model (LLM) inference is gaining significant attention. However, mobile devices operate in highly dynamic multitasking environments where users frequently switch between applications. This creates memory pressure, forc...
188. Cheap Verifiers, Large Blind Spots: Measuring the Reliability Cost of Cost-Saving Cascades ​
Author: Dushyant Rajput
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG
arXiv:2609.01345v1 Announce Type: cross Abstract: Inference cascades cut cost by answering most queries with a cheap model and escalating a hard tail to a frontier model that acts as verifier. A natural extension closes the loop: fine-tune the cheap student on the verifier's rejections so the escala...
189. Where the Verifier Fails: A Category-Level Audit of Reward Signals in RLVR ​
Author: Esther Xin
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2609.01354v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) and standard benchmark evaluation both rely on an automatic verifier that turns a free text answer into a binary reward. Prior work reports that one evaluation harness accepts only about 94% of it...
190. Exact Risk-Complexity Laws for Projective Boundaries in Scenario Optimization and Distribution-Free Certification ​
Author: Giuseppe C. Calafiore
Published: 9/2/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC
arXiv:2609.01355v1 Announce Type: cross Abstract: Scenario optimization, conformal prediction, and related distribution-free certification methods use finite samples to construct decisions or prediction sets with violation-risk guarantees for fresh observations. In several classical settings, the co...
191. Investigating Linear Probe Robustness to Linguistic Register, Medical Specialty, and Corpus Shifts in Medical QA ​
Author: Nishant Mishra, Ameen Abu-Hanna, Iacer Calixto
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2609.01361v1 Announce Type: cross Abstract: Linear classifiers trained on hidden states of a large language model (LLM), linear probes, can flag factual errors from a single forward pass. Geometrically, that implies that true and false statements separate along a stable direction in hidden sta...
192. Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity ​
Author: Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2609.01397v1 Announce Type: cross Abstract: The Rashomon effect is a machine learning phenomenon where equally accurate models produce different predictions for the same inputs (predictive multiplicity). Existing work primarily focuses on multiplicity within individual models, but in more comp...
193. On the Reliability of Generative Augmentation: A Wasserstein-Based Theoretical and Empirical Study ​
Author: Chathurika S Abeykoon, Mathias Nthiani Muia, Mallory Goldstein
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2609.01410v1 Announce Type: cross Abstract: Generative data augmentation is widely used to mitigate class imbalance, yet its theoretical effect on downstream generalization remains poorly understood. In this work, we develop a statistical framework for conditional generative augmentation and a...
194. Semantic-Guided Multimodal Preprocessing for Vision Transformer-Based Clear Cell Renal Cell Carcinoma Grading ​
Author: Fatemeh Javadian, Zhu Chen, Zahra Aminparast, Johannes Stegmaier
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV
arXiv:2609.01426v1 Announce Type: cross Abstract: Clear cell renal cell carcinoma (CCRCC) grading is essential for treatment planning, yet existing approaches either analyze patch-level images directly or focus solely on nuclei-level classification, without linking to final tumor grading. We propose...
195. Does Imitation Learning Preserve Temporal Robustness in Dexterous Manipulation? An Expert-Learner Comparison Across Task Execution Speeds ​
Author: Clinton Enwerem, John S. Baras, Calin Belta
Published: 9/2/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2609.01453v1 Announce Type: cross Abstract: Dexterous manipulation policies learned by imitation are typically evaluated for robustness to variation in scenes, objects, or instructions, but their performance across task execution speeds is less often examined. This leaves open how much tempora...
196. Sierpi\'nski--Knopp Wasserstein Distance for Persistence Diagrams and Applications to 2-Wasserstein Approximation ​
Author: Sebastien Tchitchek, Julien Tierny
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CG, cs.LG
arXiv:2609.01528v1 Announce Type: cross Abstract: This paper introduces the Sierpi'nski-Knopp (SK) Wasserstein distance, a fast metric between persistence diagrams. The SK-Wasserstein distance, denoted $d_{\mathrm{SK}}$, maps diagram points and their diagonal projections to the unit interval via th...
197. Variable Selection for Feature-Based Newsvendor ​
Author: Zhaoliang Yuan, Jie Wang
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2609.01544v1 Announce Type: cross Abstract: Feature-based newsvendor models use observable covariates to tailor inventory decisions, aiming to balance holding and shortage costs under demand uncertainty. However, high-dimensional feature sets often hinder interpretability and inflate data coll...
198. Can LLMs Discover Scientific Laws in Real and Parallel Worlds? ​
Author: Yiming Huang, Ziche Liu, Zhuohang Wu, Yiqian Wang, Junxia Cui, Xinkai Zou, Linjun Mao, Nan Huang, Naicheng Yu, Kaijie Zhu, Yue Ma, Kun Zhou, Letian Peng, Jingbo Shang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2609.01552v1 Announce Type: cross Abstract: Scientific equation discovery has long been central to scientific progress, proceeding through iterative cycles of hypothesis generation, observational testing, and refinement under scientific constraints. As LLM capabilities advance and their role i...
199. Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers ​
Author: Giovanni Bonetta, Matteo Merler, Davide Zago, Rossella Cancelliere, Bernardo Magnini
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2609.01567v2 Announce Type: cross Abstract: Vision-Language Models (VLMs) provide useful priors for interactive decision-making, but using them directly as policies is expensive and brittle: they must be queried at every step, do not improve from environment interaction, and can repeat systema...
200. Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs ​
Author: Jingtan Wang, Arun Verma, Xiaoqiang Lin, Zhengyuan Liu, Nancy F. Chen, Daniela Rus, Bryan Kian Hsiang Low
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2609.01573v1 Announce Type: cross Abstract: How to divide a fixed annotation budget between supervised fine-tuning (SFT) and reinforcement learning (RL) during LLM post-training remains an open problem. Existing work characterizes only broad trends (e.g., SFT dominates in low-data regimes), la...
201. Facet-0: A Robotic Foundation Model for Contact-Rich Precise Manipulation ​
Author: Haoyuan Deng, Haichao Liu, Wenkai Guo, Yuan Ling, Zaijia Yang, Yuanjiang Xue, Haosheng Sun, Liangzi Wang, Ziwei Wang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2609.01596v1 Announce Type: cross Abstract: Real-world robotic assembly at sub-millimeter tolerances demands spatial precision, compliant interaction, and robustness to contact failures. We present Facet-0, a robotic foundation model that predicts and values the contact consequences of its act...
202. Beyond Scores: Understanding LLM-as-a-Judge Mechanisms in Summarization Evaluation ​
Author: Himil Vasava, Ming Jiang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2609.01604v1 Announce Type: cross Abstract: LLM-based evaluators of natural language generation (NLG) quality are widely deployed as scoring tools and as automated training signals, yet the internal procedure by which they assign a rating remains poorly understood. We investigate this procedur...
203. Building Expressive and Tractable Probabilistic Generative Models: A Review ​
Author: Sahil Sidheekh, Sriraam Natarajan
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2402.00759v4 Announce Type: replace Abstract: We present a comprehensive survey of the advancements and techniques in the field of tractable probabilistic generative modeling, primarily focusing on Probabilistic Circuits (PCs). We provide a unified perspective on the inherent trade-offs betwee...
204. FedReview: Review and Dispose Poisoned Updates without Validation Datasets or Historic Knowledge ​
Author: Tianhang Zheng, Yanlu Li, Bohan Deng, Baochun Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2402.16934v2 Announce Type: replace Abstract: Federated learning has emerged as a decentralized approach for training high-performance models without accessing user data. Despite its effectiveness, it is vulnerable to poisoning attacks, where malicious users manipulate the global model by uplo...
205. Keep Everyone Happy: Online Fair Division of Numerous Items with Few Copies ​
Author: Arun Verma, Indrajit Saha, Makoto Yokoo, Bryan Kian Hsiang Low
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2408.12845v3 Announce Type: replace Abstract: This paper considers a novel variant of the online fair division problem involving multiple agents in which a learner sequentially observes an indivisible item that must be irrevocably allocated to one of the agents to achieve a desired balance bet...
206. QABBA: Symbolic Time-Series Compression via Integer-Quantized Aggregation ​
Author: Erin Carson, Xinye Chen, Fei He, Cheng Kang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, stat.ML
arXiv:2411.15209v4 Announce Type: replace Abstract: The expansion of time-series data from sensors and monitoring systems has made compact representations increasingly important. Such representations should retain signal structure while cutting storage, transmission and computation costs. Adaptive B...
207. Multi-View Causal Discovery without Non-Gaussianity: Identifiability and Algorithms ​
Author: Ambroise Heurtebise, Omar Chehab, Pierre Ablin, Alexandre Gramfort, Aapo Hyv"arinen
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2502.20115v4 Announce Type: replace Abstract: Causal discovery is a difficult problem that typically relies on strong assumptions on the data-generating model, such as non-Gaussianity. In practice, many modern applications provide multiple related views of the same system, which has rarely bee...
208. Efficient Learning of Balanced Signed Graphs via Sparse Linear Programming ​
Author: Haruki Yokota, Hiroshi Higashi, Yuichi Tanaka, Gene Cheung
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2506.01826v2 Announce Type: replace Abstract: Signed graphs are equipped with both positive and negative edge weights, encoding pairwise correlations as well as anti-correlations in data. A balanced signed graph is a signed graph with no cycles containing an odd number of negative edges. Lapla...
209. Towards Provable and Scalable Training of Quantized Neural Networks with Ising Optimization ​
Author: Wenxin Li, Chuan Wang, Hongdong Zhu, Qi Gao, Yin Ma, Hai Wei, Kai Wen
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, physics.optics
arXiv:2506.18240v5 Announce Type: replace Abstract: Training quantized neural networks remains fundamentally challenging due to non-convex loss landscapes and discrete parameter spaces. We introduce an exact Quadratic Constrained Binary Optimization (QCBO) framework with provable guarantees. We firs...
210. Any-Order GPT as Masked Diffusion Model: Decoupling Formulation and Architecture ​
Author: Shuchen Xue, Tianyu Xie, Tianyang Hu, Zijin Feng, Jiacheng Sun, Kenji Kawaguchi, Zhenguo Li, Zhi-Ming Ma
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, stat.ML
arXiv:2506.19935v2 Announce Type: replace Abstract: Efficiently scaling Large Language Models (LLMs) necessitates exploring alternatives to dominant autoregressive (AR) methods, with Masked Diffusion Models (MDMs) emerging as candidates. However, comparing AR (typically decoder-only) and MDM (often ...
211. Unsupervised Partner Design Enables Robust Ad-hoc Teamwork ​
Author: Constantin Ruhdorfer, Matteo Bortoletto, Victor Oei, Anna Penzkofer, Andreas Bulling
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC, cs.MA
arXiv:2508.06336v3 Announce Type: replace Abstract: We introduce Unsupervised Partner Design (UPD), a population-free multi-agent reinforcement learning method for robust ad-hoc teamwork. UPD generates training partners on-the-fly and selects them adaptively based on a learnability criterion, removi...
212. Recurrent State Encoders for Efficient Neural Combinatorial Optimization ​
Author: Tim Dernedde, Daniela Thyssens, Lars Schmidt-Thieme
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.05084v2 Announce Type: replace Abstract: The primary paradigm in Neural Combinatorial Optimization (NCO) consists of construction methods, where a neural network is trained to sequentially add one solution component at a time until a complete solution is formed. We observe that the typica...
213. A Compositional Kernel Model for Feature Learning ​
Author: Feng Ruan, Keli Liu, Michael Jordan
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2509.14158v3 Announce Type: replace Abstract: We study a compositional variant of kernel ridge regression in which the predictor is applied to a coordinate-wise reweighting of the inputs. Formulated as a variational problem, this model provides a tractable setting for studying feature learning...
214. Advantage Weighted Matching: Aligning RL with Pretraining in Diffusion Models ​
Author: Shuchen Xue, Chongjian Ge, Shilong Zhang, Yichen Li, Zhi-Ming Ma
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.25050v2 Announce Type: replace Abstract: Reinforcement Learning (RL) has emerged as a central paradigm for advancing Large Language Models (LLMs), where both pre-training and RL post-training stages are grounded in the same log-likelihood formulation. In contrast, recent RL approaches for...
215. Performance-Efficiency Tradeoffs in Transformers: An Approximation Theory Perspective ​
Author: Ruoxi Yu, Haotian Jiang, Jingpu Cheng, Penghao Yu, Qianxiao Li, Zhong Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2510.03784v2 Announce Type: replace Abstract: Transformers have achieved remarkable successes across a wide range of applications, yet the theoretical foundation of their model efficiency remains underexplored. In this work, we investigate how the model parameters -- mainly attention heads and...
216. Can machines think efficiently? ​
Author: Adam Winchell
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY
arXiv:2510.26954v3 Announce Type: replace Abstract: The Turing Test is no longer adequate for distinguishing human and machine intelligence. With advanced artificial intelligence systems already passing the original Turing Test and contributing to serious ethical and environmental concerns, we urgen...
217. SEBA: Sample-Efficient Black-Box Attacks on Visual Reinforcement Learning ​
Author: Tairan Huang, Yulin Jin, Junxu Liu, Qingqing Ye, Haibo Hu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2511.09681v3 Announce Type: replace Abstract: Visual reinforcement learning has achieved remarkable progress in visual control and robotics, but its vulnerability to adversarial perturbations remains underexplored. Most existing black-box attacks focus on vector-based or discrete-action RL, an...
218. Iterative GRPO: Batch-Online Policy Iteration for Multi-Turn RL via Single-Turn RLHF ​
Author: Daniel R. Jiang, Ankur Samanta, Yukai Yang, Jalaj Bhandari, R'emi Munos, Tyler Lu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.21638v2 Announce Type: replace Abstract: Practical LLM agents often operate over multi-turn conversations where success is determined only after the full interaction ends. Most multi-turn RL methods train via on-policy rollouts, but unlike in single-turn RLHF, the policy cannot produce a ...
219. Freeze, Diffuse, Decode: Task-Aware Adaptation of Transformer Embeddings for Antimicrobial Peptide Design ​
Author: Pankhil Gawade, Adam Izdebski, Myriam Lizotte, Kevin R. Moon, Jake S. Rhodes, Guy Wolf, Ewa Szczurek
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.23120v3 Announce Type: replace Abstract: Pretrained transformers provide rich, general-purpose embeddings, which are transferred to downstream tasks. However, current transfer strategies: fine-tuning and probing, either distort the pretrained geometric structure of the embeddings or lack ...
220. Training-Free Policy Violation Detection via Activation-Space Whitening in LLMs ​
Author: Oren Rachmil, Avishag Shapira, Roy Betser, Omer Hofman, Itay Gershon, Asaf Shabtai, Yuval Elovici, Roman Vainshtein
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.03994v4 Announce Type: replace Abstract: As organizations increasingly deploy LLMs in sensitive domains such as legal, financial, and medical settings, ensuring alignment with internal organizational policies has become a priority. Existing content moderation frameworks remain largely con...
221. Control Variate Score Matching for Diffusion Models ​
Author: Khaled Kahouli, Romuald Elie, Klaus-Robert M"uller, Quentin Berthet, Oliver T. Unke, Arnaud Doucet
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.20003v2 Announce Type: replace Abstract: Sampling from unnormalized probability densities is a pervasive challenge across the computational and physical sciences. Diffusion models provide a powerful generative framework for this task, but their success relies on accurately estimating the ...
222. Variance-Adaptive Muon: Pre-Orthogonalization Variance Modulation for Efficient Language Model Pretraining ​
Author: Jingru Li, Yibo Fan, Huan Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.14603v2 Announce Type: replace Abstract: Optimizer design plays a central role in efficient language model pretraining, directly affecting optimization dynamics, convergence speed, and compute cost under fixed training budgets. Muon has emerged as a strong optimizer by orthogonalizing mom...
223. FloydNet: A Learning Paradigm for Global Relational Reasoning ​
Author: Jingcheng Yu, Mingliang Zeng, Qiwei Ye
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2601.19094v3 Announce Type: replace Abstract: Learning algorithmic computation often requires explicit relational intermediate states, yet many graph processors maintain their primary states on individual entities. We introduce \fnet and \textbf{Pivotal Attention} (PA), which maintain ordered ...
224. Breaking the Reasoning Horizon in Entity Alignment Foundation Models ​
Author: Yuanning Cui, Zequn Sun, Wei Hu, Kexuan Xin, Zhangjie Fu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.21174v3 Announce Type: replace Abstract: Entity alignment (EA) is critical for knowledge graph (KG) fusion. Existing EA models lack transferability and are incapable of aligning unseen KGs without retraining. While using graph foundation models (GFMs) offer a solution, we find that direct...
225. Learning to Remember: End-to-End Training of Memory Agents for Long-Context Reasoning ​
Author: Kehao Zhang, Shangtong Gui, Sheng Yang, Wei Chen, Yang Feng
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.18493v2 Announce Type: replace Abstract: Long-context LLMs and Retrieval-Augmented Generation defer state tracking and evidence consolidation to query time, which is brittle when facts evolve and answers depend on latent states. We introduce Unified Memory Agent (UMA) for a one-to-many se...
226. Efficient Adaptation of ROMs for Unsteady Flows Using Data Assimilation ​
Author: Isma"el Zighed, Andrea N'ovoa, Luca Magri, Taraneh Sayadi
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn
arXiv:2602.23188v3 Announce Type: replace Abstract: We propose an efficient retraining strategy for a parameterized Reduced Order Model (ROM) that attains accuracy comparable to full retraining while requiring only a fraction of the computational time and relying solely on sparse observations of the...
227. Inverse Reconstruction of Shock Time Series from Shock Response Spectrum Curves using Machine Learning ​
Author: Adam Watts (Los Alamos National Laboratory), Andrew Jeon (Los Alamos National Laboratory), Destry Newton (Los Alamos National Laboratory), Ryan Bowering (University of Rochester)
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2603.03229v3 Announce Type: replace Abstract: The shock response spectrum (SRS) is widely used to characterize the response of single-degree-of-freedom (SDOF) systems to transient accelerations. Because the mapping from acceleration time history to SRS is nonlinear and many-to-one, reconstruct...
228. MMAI Gym for Science: Training Liquid Foundation Models for Drug Discovery ​
Author: Maksim Kuznetsov, Zulfat Miftahutdinov, Rim Shayakhmetov, Mikolaj Mizera, Roman Schutski, Bogdan Zagribelnyy, Ivan Ilin, Nikita Bondarev, Thomas MacDougall, Mathieu Reymond, Mihir Bafna, Kaeli Kaymak-Loveless, Eugene Babin, Maxim Malkov, Mathias Lechner, Ramin Hasani, Alexander Amini, Vladimir Aladinskiy, Alex Aliper, Alex Zhavoronkov
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2603.03517v2 Announce Type: replace Abstract: General-purpose large language models (LLMs) that rely on in-context learning do not reliably deliver the scientific understanding and performance required for drug discovery tasks. Simply increasing model size or introducing reasoning tokens does ...
229. RetroReasoner: A Reasoning LLM for Strategic Retrosynthesis Prediction ​
Author: Hanbum Ko, Chanhui Lee, Ye Rin Kim, Rodrigo Hormazabal, Sehui Han, Sungbin Lim, Sungwoong Kim
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.12666v3 Announce Type: replace Abstract: Retrosynthesis prediction aims to identify reactants that can synthesize a given product molecule. Although molecular large language models (LLMs) have recently shown promising results, most existing methods either generate reactants directly or pr...
230. SCALE:Scalable Conditional Atlas-Level Endpoint transport for virtual cell perturbation prediction ​
Author: Shuizhou Chen, Lang Yu, Xueqin Lin, Xinjie Mao, Songming Zhang, Xinyu Gu, Hao Wu, Sheng Xu, Kedu Jin, Lei Bai, Quan Qian, Qin Chen, Qiang Gao, Siqi Sun, Zhangyang Gao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.QM
arXiv:2603.17380v3 Announce Type: replace Abstract: Virtual-cell models aim to predict how cell populations respond to perturbations, but control and treated cells are measured as unpaired populations, complicating the learning of perturbation-specific effects. We present SCALE, a conditional transp...
231. Uniform a priori bounds and error analysis for the Adam stochastic gradient descent optimization method ​
Author: Steffen Dereich, Thang Do, Arnulf Jentzen
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2603.18899v2 Announce Type: replace Abstract: The adaptive moment estimation (Adam) optimizer proposed by Kingma & Ba (2014) is presumably the most popular stochastic gradient descent (SGD) optimization method for the training of deep neural networks (DNNs) in artificial intelligence (AI) syst...
232. Rigorous Error Certification for Neural PDE Solvers: From Empirical Residuals to Solution Guarantees ​
Author: Amartya Mukherjee, Maxwell Fitzsimmons, David C. Del Rey Fern'andez, Jun Liu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, math.AP, math.FA
arXiv:2603.19165v2 Announce Type: replace Abstract: Uncertainty quantification for partial differential equations is traditionally grounded in discretization theory, where solution error is controlled via mesh/grid refinement. Physics-informed neural networks fundamentally depart from this paradigm:...
233. Process-Aware AI for Rainfall-Runoff Modeling: A Mass-Conserving Neural Framework with Hydrological Process Constraints ​
Author: Mohammad A. Farmani, Hoshin V. Gupta, Ali Behrangi, Muhammad Jawad, Sadaf Moghisi, Guo-Yue Niu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.25093v2 Announce Type: replace Abstract: Machine learning models can achieve high predictive accuracy in hydrological applications but often lack physical interpretability. The Mass-Conserving Perceptron (MCP) provides a physics-aware artificial intelligence (AI) framework that enforces c...
234. KV Cache Offloading for Context-Intensive Tasks ​
Author: Andrey Bocharnikov, Ivan Ermakov, Denis Kuznedelev, Vyacheslav Zhdanovskiy, Yegor Yershov
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2604.08426v5 Announce Type: replace Abstract: With the growing demand for long-context LLMs across a wide range of applications, the key-value (KV) cache has become a critical bottleneck for both latency and memory usage. Recently, KV-cache offloading has emerged as a promising approach to red...
235. What Drives Representation Steering? A Mechanistic Case Study on Steering Refusal ​
Author: Stephen Cheng, Sarah Wiegreffe, Dinesh Manocha
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2604.08524v2 Announce Type: replace Abstract: Applying steering vectors to large language models (LLMs) is an efficient and effective model alignment technique, but we lack an interpretable explanation for how it works--specifically, what internal mechanisms steering vectors affect and how thi...
236. Global Attention with Linear Complexity for Exascale Generative Data Assimilation in Earth System Prediction ​
Author: Xiao Wang, Zezhong Zhang, Isaac Lyngaas, Hong-Jun Yoon, Jong-Youl Choi, Siming Liang, Janet Wang, Hristo G. Chipilski, Ashwin M. Aji, Feng Bao, Peter Jan van Leeuwen, Dan Lu, Guannan Zhang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2604.16590v2 Announce Type: replace Abstract: Accurate Earth system prediction requires state inference from incomplete observations, but conventional two-stage data assimilation (DA) is computationally prohibitive because repeated PDE-based ensemble forecasts, observation updates, and interme...
237. The Topological Trouble With Transformers ​
Author: Michael C. Mozer, Shoaib Ahmed Siddiqui, Rosanne Liu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2604.17121v5 Announce Type: replace Abstract: Transformers encode structure in sequences via an expanding contextual history. However, their purely feedforward architecture fundamentally limits dynamic state tracking. State tracking -- the iterative updating of latent variables reflecting an e...
238. Reparameterization through Coverings and Topological Weight Priors ​
Author: Maxim Beketov, Pavel Snopov
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.23804v2 Announce Type: replace Abstract: We generalise the reparameterization trick (RT) applied in variational autoencoders (VAEs) letting these have latent spaces of non-trivial topology - i.e. that of base manifolds covered with other ones, on which some technique for RT is available. ...
239. Universal Approximation of Nonlinear Operators and Their Derivatives ​
Author: Filippo de Feo
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NA, math.FA, math.NA, math.OC
arXiv:2605.15285v3 Announce Type: replace Abstract: Establishing Universal Approximation Theorems (UATs) for nonlinear operators and their derivatives is a foundational open problem in Operator Learning (OL) and raises delicate questions in Nonlinear Functional Analysis. We prove the first UATs for ...
240. Why Do Reasoning Models Lose Coverage? The Role of Data and Forks in the Road ​
Author: Ngoc-Hieu Nguyen, Parshin Shojaee, Phuc Minh Nguyen, Nan Zhang, Chandan K Reddy, Khoa D Doan, Rui Zhang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.17026v2 Announce Type: replace Abstract: Recent progress in large language models has led to the emergence of reasoning models, which have shown strong performance on complex tasks through specialized fine-tuning procedures. While these methods reliably improve pass@1 accuracy, prior work...
241. LLM-driven design of physics-constrained constitutive models: two agents are better than one ​
Author: Marius Tacke, Matthias Busch, Kian Abdolazizi, Jonas Eichinger, Kevin Linka, Roland Aydin, Christian Cyron
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.23754v2 Announce Type: replace Abstract: Developing constitutive models that capture how materials deform under load traditionally requires years of specialized expertise in continuum mechanics, machine learning, and scientific programming. Large language models (LLMs) have recently been ...
242. Latent Recurrent Transformer: Architecture Exploration, Training Strategies, and Scaling Behavior ​
Author: Zeyi Huang, Xuehai He, LiLiang Ren, Yiping Wang, Baolin Peng, Hao Cheng, Shuohang Wang, Pengcheng He, Jianfeng Gao, Yong Jae Lee, Yelong Shen
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2605.26797v2 Announce Type: replace Abstract: We study Latent Recurrent Transformer (LRT), a lightweight augmentation of autoregressive transformers that reuses a high-level source-layer hidden state from the previous token as recurrent memory for the next token. Because this state is already ...
243. When the Strongest Teacher Is Not the Best Teacher: Student-Centric Answer Selection ​
Author: Zhengyu Hu, Zheyuan Xiao, Linxin Song, Fengqing Jiang, Yuetai Li, Zhihan Xiong, Yue Liu, Junhao Lin, Yao Su, Lijie Hu, Kaize Ding, Teng Xiao, Radha Poovendran
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2605.26872v5 Announce Type: replace Abstract: LLM training increasingly relies on teacher-generated supervision, from synthetic responses to reasoning traces and tool-use demonstrations. Current practice often chooses the highest-performing teacher to generate student training data, implicitly...
244. PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning ​
Author: Qikai Chang, Zhenrong Zhang, Linbo Chen, Pengfei Hu, Jianshu Zhang, Youhui Guo, Jun Du
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2605.29582v2 Announce Type: replace Abstract: Large Language Models (LLMs) show strong potential as educational tutors. Existing approaches typically train them to solve problems and provide correct answers, but this problem-solving-centered paradigm overlooks key requirements of effective tut...
245. Skill Reuse as Compression in Agentic RL ​
Author: Zhikun Xu, Yu Feng, Jacob Dineen, Taiwei Shi, Jieyu Zhao, Ben Zhou
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.31509v2 Announce Type: replace Abstract: Large language model agents trained with reinforcement learning (RL) often learn brittle, task-specific shortcuts. We hypothesize that agents generalize better when their successful trajectories are structurally compressible, decomposed into a smal...
246. What Do Students Learn? A Feature-Level Analysis of Dark Knowledge ​
Author: Seungu Kang, Songkuk Kim
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.03052v2 Announce Type: replace Abstract: Knowledge Distillation (KD) is a powerful tool for model compression, yet the precise mechanisms by which student models acquire feature representations remain underexplored. In this work, we analyze student feature learning using the Interaction T...
247. RECAP: Regression Evaluation for Continual Adaptation of Prompts ​
Author: Harsh Deshpande, Kushal Chawla, Sangwoo Cho, William Campbell, Sambit Sahu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2606.06698v4 Announce Type: replace Abstract: Production agentic systems routinely face evolving constraints and must comply from the very next interaction. Scenarios like a tool-call notification changing a compliance threshold or a policy update adding disclosure requirements fit this criter...
248. Enabling KV Caching of Shared Prefix for Diffusion Language Models ​
Author: Younghun Go, Jaehoon Han, Changyong Shin, Chuck Yoo, Gyeongsik Yang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.07571v4 Announce Type: replace Abstract: Key-value (KV) caching for shared prefixes is essential for high-throughput large language model (LLM) serving, but it faces critical challenges in emerging diffusion language models (DLMs). In DLMs, bidirectional attention means that updating any ...
249. DOG-DPO:Dynamic Optimization in Geometry for Safety Alignment ​
Author: Yi Nian, Tiankai Yang, Yudi Zhang, Qi Pan, Zelong Xu, Shenzhe Zhu, Qingqing Luan, Yue Huang, Xiangliang Zhang, Yue Zhao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.07678v4 Announce Type: replace Abstract: Safety alignment for large language models relies on preference data, but current pipelines often train on large, redundant datasets. Existing data selection methods typically score each preference pair independently, collapsing directional prefere...
250. Learning to Refine Hidden States for Reliable LLM Reasoning ​
Author: Chia-Hsuan Hsu, Jui-Ming Yao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.17524v3 Announce Type: replace Abstract: Large language models show strong reasoning ability, but their internal reasoning process can remain unstable in complex multi-step settings, where early hidden-state errors may propagate to incorrect predictions. We propose ReLAR, a reinforcement-...
251. Beyond AHI: An Interpretable Causal-Discovery-Guided Framework for Sleep Recovery in Connected Health ​
Author: Saba A. Farahani, Elahe Khatibi, Manoj Vishwanath, Amir M. Rahmani, Hung Cao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, stat.AP
arXiv:2606.18506v2 Announce Type: replace Abstract: Objective sleep assessment relies on polysomnography (PSG), yet clinical impact is often better reflected in patient-reported outcomes (PROs) such as sleepiness and fatigue. Existing summary indices, including the Apnea-Hypopnea Index (AHI), provid...
252. Final Checkpoints Are Not Enough: Analyzing Latent Reasoning Faithfulness Along Training Trajectories ​
Author: Hengyu Jin, Shu Yang, Di Wang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.06648v2 Announce Type: replace Abstract: Latent reasoning performs multi-step inference in continuous hidden states, promising more compact and efficient reasoning. However, these opaque states raise a question of faithfulness: whether the latent reasoning steps drive the final answer. Pr...
253. Shallower ReLU Network Representations via Exact Linear Algebra ​
Author: Kilian Rue{\ss}, Gennadiy Averkov, Florestan Brunck, Moritz Grillo, Christoph Hertrich, Georg Loho, Jack Stade, Moritz Stargalla, Matthew Sun, Martin Winter
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, math.CO
arXiv:2607.21651v2 Announce Type: replace Abstract: We study the depth required by ReLU networks to exactly represent piecewise linear functions, focusing specifically on the maximum function. This problem has recently received significant attention in both the ML and TCS literature. We prove that $...
254. S-CEReBrO: Breaking the Memory Barrier in Continuous EEG Monitoring ​
Author: Glenn Anta Bucagu, Thorir Mar Ingolfsson, Yawei Li, Luca Benini
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27913v2 Announce Type: replace Abstract: Foundation models offer a promising paradigm for Electroencephalography (EEG) analysis, leveraging generalizable representations from vast unlabeled datasets. Yet, Transformer-based architectures face a critical bottleneck: global attention mechani...
255. Can We Trust In-Distribution Success? Locked Evaluation Reveals Transfer Failure and Sampling-Depth Entanglement in CRISPRi Perturbation Prediction ​
Author: Mehrdad Shoeibi, Niloofar Yousefi
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.00152v3 Announce Type: replace Abstract: AI evaluation can support the wrong inference when an in-domain benchmark success does not survive distribution shift, or when the benchmark endpoint is entangled with a design factor. We study this problem in CRISPRi perturbation-effect prediction...
256. Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility ​
Author: Mohsen Hariri, Weicong Chen, Nahal Shahini, Vikash Singh, Kai Ye, Amirhossein Samandar, Debargha Ganguly, Sreehari Sankar, Yanyan Zhang, Shouren Wang, Jerry Peng, Biyao Zhang, Michael Hinczewski, Vipin Chaudhary
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.04001v2 Announce Type: replace Abstract: Large language models can solve harder reasoning problems with more inference-time compute. The term "test-time scaling," however, covers several inference algorithms: extending deliberation along one trajectory, sampling completed candidates and a...
257. Reading the Gate, Not the Interference: Output-Side Interference Measurement Does Not Track Merge Collapse ​
Author: Chencheng Zhu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11797v3 Announce Type: replace Abstract: Task-arithmetic merging works until it doesn't, and the field diagnoses why by measuring interference inside the merged model. We take the most direct such measure, the exact layerwise activation cross-term of a factorial ledger, establish its caus...
258. Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling ​
Author: Michael Fore, Akshay Jain, Justin Downes, Rohan Pradhan, Duncan Botti
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14349v2 Announce Type: replace Abstract: We present a training-free method for multi-modal trajectory prediction that achieves comparable accuracy to a 57M-parameter transformer while requiring no GPU and zero learned parameters. The method builds a transition table of historical state-to...
259. Beyond Dense Adam States: Adaptive Log-Space Quantization for Memory-Efficient Optimizers ​
Author: Yan Wang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.22322v3 Announce Type: replace Abstract: Optimizer-state quantization is commonly designed for Adam's dense, parameter-aligned first- and second-moment arrays. This abstraction breaks for memory-efficient optimizers, whose states may be factored, confidence-modulated, or maintained in a p...
260. Stress Testing Unlearning Algorithms ​
Author: Noam Diamant, Neta Glazer, Ethan Fetaya
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.22527v2 Announce Type: replace Abstract: Recently, machine unlearning, the removal of specific training data influence from a model, has gained increasing attention. In large language models (LLMs), unlearning is particularly challenging due to the ambiguity of inputs and outputs. Con- se...
261. The Frame Kernel Method for Multiscale Operator Learning ​
Author: Branden Frieden, Ryan Whitehead, M. Keith Ballard, Robert M. Kirby, Varun Shankar
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.25084v2 Announce Type: replace Abstract: We present a natively multiscale operator learning method for the surrogate modeling of (numerical solvers for) multiscale partial differential equations (PDEs). The primary novelty of our method lies in a novel multiscale kernel frame function app...
262. Performative Privacy: When Differential Privacy Maximizes Utility ​
Author: Uddalak Mukherjee, Edwige Cyffers, Yann Chevaleyre
Published: 9/2/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.28198v2 Announce Type: replace Abstract: Privacy-preserving learning is often motivated by the idea that protecting users' data can preserve trust and thus participation, improving utility in the long term. However, this claim has not been formalized so far. In parallel, performative lear...
263. Deep learning based numerical approximation algorithms for stochastic partial differential equations ​
Author: Christian Beck, Sebastian Becker, Patrick Cheridito, Arnulf Jentzen, Ariel Neufeld
Published: 9/2/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.PR, stat.ML
arXiv:2012.01194v3 Announce Type: replace-cross Abstract: In this article, we introduce a deep learning based approximation algorithm for SPDEs. Our approach employs neural networks to approximate the solutions of SPDEs along given realizations of the driving noise process. If applied to a set of si...
264. GENIE: Watermarking Graph Neural Networks for Link Prediction ​
Author: Venkata Sai Pranav Bachina, Aaryan Ajay Sharma, Ankit Gangwal, Charu Sharma
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2406.04805v4 Announce Type: replace-cross Abstract: The rapid adoption, usefulness, and resource-intensive training of Graph Neural Network (GNN) models have made them an invaluable intellectual property in graph-based machine learning. However, their wide-spread adoption also makes them susce...
265. Generalization Bounds for Markov Algorithms through Entropy Flow Computations ​
Author: Benjamin Dupuis, Maxime Haddouche, George Deligiannidis, Umut Simsekli
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2502.07584v3 Announce Type: replace-cross Abstract: Many learning algorithms can be represented as Markov processes, and understanding their generalization error is a central topic in learning theory. For specific continuous-time noisy algorithms, a prominent analysis technique relies on infor...
266. Online simultaneous inference for quantiles via smoothed stochastic gradient descent ​
Author: Likai Chen, Georg Keilbar, Wei Biao Wu
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2505.13299v2 Announce Type: replace-cross Abstract: This paper considers the estimation of quantiles via a smoothed version of the stochastic gradient descent (SGD) algorithm. By smoothing the score function with a bandwidth tied to the learning rate, we obtain estimates that are monotone in t...
267. On the Existence of Consistent Adversarial Attacks in High-Dimensional Linear Classification ​
Author: Matteo Vilucchio, Lenka Zdeborov'a, Bruno Loureiro
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.dis-nn, cs.CR, cs.LG
arXiv:2506.12454v2 Announce Type: replace-cross Abstract: What fundamentally distinguishes an adversarial attack from a misclassification due to limited model expressivity or finite data? In this work, we investigate this question in the setting of high-dimensional binary classification, where stati...
268. BiasGym: A Simple and Generalizable Framework for Analyzing and Removing Biases through Injection ​
Author: Sekh Mainul Islam, Nadav Borenstein, Siddhesh Milind Pawar, Haeun Yu, Arnav Arora, Isabelle Augenstein
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2508.08855v5 Announce Type: replace-cross Abstract: Understanding biases and stereotypes encoded in the weights of Large Language Models (LLMs) is crucial for developing effective mitigation strategies. However, biased behavior is often subtle and non-trivial to isolate, even when deliberately...
269. FlexP-SFT: A Flexible Aggregation-Free Framework for On-Device Personalized Split Federated Fine-Tuning of LLMs ​
Author: Jiaxiang Geng, Tianjun Yuan, Pengchao Han, Ying Gao, Xianhao Chen, Bing Luo
Published: 9/2/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2508.10349v2 Announce Type: replace-cross Abstract: To fine-tune large language models (LLMs) over private data, federated learning (FL) has emerged as a promising paradigm. However, the prohibitive memory and communication demands of LLMs render standard FL impractical for resource-constraine...
270. SupraTok: Cross-Boundary Tokenization for Enhanced Language Model Performance ​
Author: Andrei-Valentin T\u{a}nase, Elena Pelican
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2508.11857v3 Announce Type: replace-cross Abstract: Tokenization remains a persistent bottleneck in language modeling, especially when vocabulary learning is limited by whitespace boundaries. We present SupraTok, a tokenizer that crosses whitespace boundaries using three modular components: op...
271. Integrated Noise and Safety Management in UAM via A Unified Reinforcement Learning Framework ​
Author: Surya Murthy, Zhenyu Gao, John-Paul Clarke, Ufuk Topcu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2508.16440v3 Announce Type: replace-cross Abstract: Urban Air Mobility (UAM) envisions the widespread use of small aerial vehicles to transform transportation in dense urban environments. However, UAM faces critical operational challenges, particularly the balance between minimizing noise expo...
272. Fair Minimum Labeling: Efficient Temporal Network Activations for Reachability and Equity ​
Author: Lutz Oettershagen, Othon Michail
Published: 9/2/2026, 4:00:00 AM
Categories: cs.SI, cs.DS, cs.LG
arXiv:2510.03899v3 Announce Type: replace-cross Abstract: Balancing resource efficiency and fairness is critical in networked systems that support modern learning applications. We introduce the \emph{Fair Minimum Labeling} (FML) problem: the task of designing a minimum-cost temporal edge activation ...
273. Silence is Golden: Mitigating Hallucinations in Large Audio-Language Models via Layer-Weighted Vector Steering ​
Author: Tsung-En Lin, Kuan-Yi Lee, Hung-Yi Lee
Published: 9/2/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS
arXiv:2510.12851v2 Announce Type: replace-cross Abstract: Large Audio-Language Models (LALMs) excel in Audio QA but often suffer from hallucinations ungrounded in the audio. To our knowledge, we are the first to propose applying vector steering to the audio domain to mitigate this. Unlike text-based...
274. Compositional Machine Design as Program Synthesis with LLMs ​
Author: Wenqian Zhang, Yangyi Huang, Weiyang Liu, Zhen Liu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CV, cs.GR, cs.LG
arXiv:2510.14980v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong abilities in writing and revising programs, yet many program-synthesis benchmarks still evaluate programs in symbolic or digital environments. We introduce compositional machine design, a physica...
275. Nonlinear Dynamics In Optimization Landscape of Shallow Neural Networks with Tunable Leaky ReLU ​
Author: Jingzhou Liu
Published: 9/2/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DS
arXiv:2510.25060v2 Announce Type: replace-cross Abstract: In this work, we study the nonlinear dynamics of a shallow neural network trained with mean-squared loss and leaky ReLU activation. Under Gaussian inputs and equal layer width k, (1) we establish, based on the equivariant gradient degree, a t...
276. Multi-Step Knowledge Interaction Analysis via Rank-2 Subspace Disentanglement ​
Author: Sekh Mainul Islam, Pepa Atanasova, Isabelle Augenstein
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2511.01706v3 Announce Type: replace-cross Abstract: Natural Language Explanations (NLEs) describe how Large Language Models (LLMs) make decisions by drawing on external Context Knowledge (CK) and Parametric Knowledge (PK). Understanding the interaction between these sources is key to assessing...
277. The Alexander-Hirschowitz theorem for neurovarieties ​
Author: A. Massarenti, M. Mella
Published: 9/2/2026, 4:00:00 AM
Categories: math.AG, cs.AI, cs.LG, math.AC
arXiv:2511.19703v2 Announce Type: replace-cross Abstract: We study the dimension and identifiability of neurovarieties associated to polynomial neural networks. We give an independent geometric proof that the linear bounds $d_i\geq 2n_i-1$ on the activation degrees imply non defectiveness for any nu...
278. 3D-Consistent Multi-View Editing by Correspondence Guidance ​
Author: Josef Bengtson, David Nilsson, Dong In Lee, Yaroslava Lochman, Fredrik Kahl
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2511.22228v3 Announce Type: replace-cross Abstract: Recent advancements in diffusion and flow models have greatly improved text-based image editing, yet methods that edit images independently often produce geometrically and photometrically inconsistent results across different views of the sam...
279. Probabilistic Multi-Agent Aircraft Landing Time Prediction ​
Author: Kyungmin Kim, Seokbin Yoon, Keumjin Lee
Published: 9/2/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2512.08281v2 Announce Type: replace-cross Abstract: Accurate and reliable aircraft landing time prediction is essential for effective resource allocation in air traffic management. However, the inherent uncertainty of aircraft trajectories and traffic flows poses significant challenges to both...
280. CADKnitter: Compositional CAD Generation from Text and Geometry Guidance ​
Author: Tri Le, Khang Nguyen, Baoru Huang, Tung D. Ta, Anh Nguyen
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2512.11199v2 Announce Type: replace-cross Abstract: Computer-aided design (CAD) defines 3D models as compact, precise, and editable representations, making it directly useful for several fields. Recently, CAD generation has been gaining more attention in both the research community and industr...
281. Modeling Information Blackouts in Missing Not-At-Random Time Series Data ​
Author: Aman Sunesh (New York University), Allan Ma (New York University), Siddarth Nilol (New York University)
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP
arXiv:2601.01480v3 Announce Type: replace-cross Abstract: Traffic forecasting systems rely on fixed sensor networks that frequently exhibit contiguous blackouts. Such outages are usually treated as ignorable missingness, although dropout can depend on unobserved traffic conditions. We study this pos...
282. Hidden State Poisoning Attacks against Mamba-based Language Models ​
Author: Alexandre Le Mercier, Chris Develder, Thomas Demeester
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2601.01972v5 Announce Type: replace-cross Abstract: State space models (SSMs) like Mamba offer efficient alternatives to Transformer-based language models, with linear time complexity. Yet, their adversarial robustness remains critically unexplored. This paper studies the phenomenon whereby sp...
283. DAGGER: Distractor-Aware Graph Generation for Executable Reasoning in Math Problems ​
Author: Zabir Al Nazi, Shubhashis Roy Dipta, Sudipta Kar
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2601.06853v3 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) prompting is widely adopted for mathematical problem solving, including in low-resource languages, yet its behavior under irrelevant context remains underexplored. To systematically study this challenge, we introduce DI...
284. Auditing Frozen-Encoder Anomaly Detection Across Mechanical Systems: Representation Provenance, Calibration, and Protocol Effects ​
Author: Jose S'anchez Andreu
Published: 9/2/2026, 4:00:00 AM
Categories: astro-ph.IM, cs.LG, physics.data-an
arXiv:2601.11415v2 Announce Type: replace-cross Abstract: This version reports a reproducibility audit of the frozen-encoder experiments presented in version 1. The numerical discrimination results are reproducible from the preserved artifacts, but their original attribution to interferometric pretr...
285. Online Regime-aware Calibration for Black-box Social Simulators via Posterior-assisted Evolutionary Dynamic Optimization ​
Author: Peng Yang, Zhenhua Yang, Boquan Jiang, Chenkai Wang, Ke Tang, Xin Yao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.NE, cs.LG
arXiv:2601.19481v2 Announce Type: replace-cross Abstract: Evolutionary dynamic optimization (EDO) commonly assumes that environmental changes can be detected from fitness variations and handled through random re-initialization, historical solutions, or learned transition patterns. Online calibration...
286. Denoising the Deep Sky: Physics-Based CCD Noise Formation for Astronomical Imaging ​
Author: Shuhong Liu, Xining Ge, Ziying Gu, Quanfeng Xu, Lin Gu, Ziteng Cui, Xuangeng Chu, Jun Liu, Dong Li, Tatsuya Harada
Published: 9/2/2026, 4:00:00 AM
Categories: astro-ph.IM, cs.CV, cs.LG
arXiv:2601.23276v4 Announce Type: replace-cross Abstract: Astronomical imaging remains noise-limited under practical observing conditions. Standard calibration pipelines remove structured artifacts but largely leave stochastic noise unresolved. Although learning-based denoising has shown strong pote...
287. Persistent Entropy as a Detector of Phase Transitions ​
Author: Marcos Gutierrez-del-Pozo, Eduardo Paluzo-Hidalgo, Matteo Rucco
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.IT, cs.LG, math.IT
arXiv:2602.09058v2 Announce Type: replace-cross Abstract: Persistent entropy is a scalar summary of persistence barcodes widely used to detect regime changes, yet there is no account of when a structural change in a barcode must produce a detectable change in entropy. We establish a model-agnostic t...
288. Is Knowledge Distillation Actually Greener? A Case Study in Machine Translation ​
Author: Joseph Attieh, Timothee Mickus, Anne-Laure Ligozat, Aur'elie N'ev'eol, J"org Tiedemann
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2602.09691v2 Announce Type: replace-cross Abstract: Knowledge distillation (KD) is a technique to compress a larger teacher system into a smaller student. In machine translation, KD is commonly evaluated through translation quality and inference efficiency, without jointly accounting for the e...
289. Ontology-Guided Neuro-Symbolic Inference: Grounding Language Models with Mathematical Domain Knowledge ​
Author: Marcelo Labre
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SC
arXiv:2602.17826v2 Announce Type: replace-cross Abstract: Language models exhibit fundamental limitations -- hallucination, brittleness, and lack of formal grounding -- that are particularly problematic in high-stakes specialist fields requiring verifiable reasoning. I investigate whether formal dom...
290. Channel-Adaptive Edge AI: Maximizing Inference Throughput by Adapting Computational Complexity to Channel States ​
Author: Jierui Zhang, Jianhao Huang, Kaibin Huang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, cs.NI, math.IT
arXiv:2603.03146v2 Announce Type: replace-cross Abstract: \emph{Integrated communication and computation} (IC$^2$) has emerged as a new paradigm for enabling efficient edge inference in sixth-generation (6G) networks. However, the design of IC$^2$ technologies is hindered by the lack of a tractable ...
291. HEAL: Hindsight Entropy-Assisted Learning for Reasoning Distillation ​
Author: Wenjing Zhang, Jiangze Yan, Jieyun Huang, Yi Shen, Shuming Shi, Ping Chen, Ning Wang, Zhaoxiang Liu, Kai Wang, Shiguo Lian
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2603.10359v2 Announce Type: replace-cross Abstract: Distilling reasoning capabilities from Large Reasoning Models (LRMs) into smaller models is typically constrained by the limitations of rejection sampling. Standard methods treat the teacher as a static filter, discarding complex "corner-case...
292. MineDraft: A Framework for Batch Parallel Speculative Decoding ​
Author: Zhenwei Tang, Arun Verma, Zijian Zhou, Zhaoxuan Wu, Alok Prakash, Daniela Rus, Bryan Kian Hsiang Low
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.DC, cs.LG
arXiv:2603.18016v3 Announce Type: replace-cross Abstract: Speculative decoding (SD) accelerates large language model inference by using a smaller draft model to propose draft tokens that are subsequently verified by a larger target model. However, the performance of standard SD is often limited by t...
293. A penalised Saito functional for heuristic search of free line arrangements ​
Author: Tom'as S. R. Silva
Published: 9/2/2026, 4:00:00 AM
Categories: math.AG, cs.LG, math.CO
arXiv:2604.02995v3 Announce Type: replace-cross Abstract: We introduce the penalised Saito functional $\mathfrak S_{\lambda,\beta}(\mathcal{A};d_1,d_2)$ for a reduced arrangement $\mathcal{A}$ of $n$ lines and a prescribed pair $d_1+d_2=n-1$. It measures the alignment of a candidate Saito determinan...
294. Why Fine-Tuning Encourages Hallucinations and How to Fix It ​
Author: Guy Kaplan, Zorik Gekhman, Zhen Zhu, Lotem Rozner, Yuval Reif, Swabha Swayamdipta, Derek Hoiem, Roy Schwartz
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.NE
arXiv:2604.15574v2 Announce Type: replace-cross Abstract: Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information through supervised fine-tuning (SFT), which can increase hallucinations w.r.t.~knowledge acqu...
295. FedSPDnet: Geometry-Aware Federated Deep Learning with SPDnet ​
Author: Thibault Pautrel, Florent Bouchard, Ammar Mian, Guillaume Ginolhac
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2604.22494v2 Announce Type: replace-cross Abstract: We introduce two federated learning frameworks for the classical SPDnet model operating on symmetric positive definite (SPD) matrices with Stiefel-constrained parameters. Unlike standard Euclidean averaging, which violates orthogonality, our ...
296. D3-Gym: Constructing Real-World Verifiable Environments for Data-Driven Discovery ​
Author: Hanane Nour Moussa, Yifei Li, Zhuoyang Li, Yankai Yang, Cheng Tang, Tianshu Zhang, Nesreen K. Ahmed, Ali Payani, Ziru Chen, Huan Sun
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2604.27977v4 Announce Type: replace-cross Abstract: Despite recent progress in language models and agents for scientific data-driven discovery, advancing their capabilities is held back by the absence of verifiable environments representing real-world scientific tasks. To fill this gap, we int...
297. Leakage-Audited Benchmarking Reveals Limited Evidence for Cross-Subject Auditory-Evoked EEG Vowel Perception Decoding ​
Author: Xiaoyang Li, Zeyan Tao
Published: 9/2/2026, 4:00:00 AM
Categories: eess.SP, cs.CL, cs.CV, cs.LG, cs.SD, q-bio.NC
arXiv:2605.00865v4 Announce Type: replace-cross Abstract: We tested whether auditory-evoked EEG supports subject-independent five-vowel perception decoding when trial identity, model identity, prediction provenance, and participant-level inference are controlled within a single benchmark. We reconst...
298. Polarizable atomic multipoles for learning long-range electrostatics ​
Author: Yoonjae Park, Dongjin Kim, Daniel S. King, Nam H. {\DJ}`ao, Roya Savoj, Sebastien Hamel, Xiaoyu Wang, Bingqing Cheng
Published: 9/2/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG, physics.chem-ph, physics.comp-ph
arXiv:2605.05746v2 Announce Type: replace-cross Abstract: Long-range electrostatics and polarization remain central obstacles to extending machine learning interatomic potentials (MLIPs) to ionic, polar, and interfacial systems. Here we introduce a semi-local framework for learning electrostatics fr...
299. DiscoverPhysics: Benchmarking LLMs for Out-of-the-Box Scientific Thinking ​
Author: Matt L. Wiemann, Lindsay M. Smith, Peter Melchior, Siddharth Mishra-Sharma, Andrew Gordon Wilson, Pavel Izmailov, Carolina Cuesta-L'azaro
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2605.26087v2 Announce Type: replace-cross Abstract: Frontier LLMs now perform strongly across a wide range of physics evaluations, but it is hard to disentangle genuine reasoning from recall of established science. We introduce DiscoverPhysics, an interactive benchmark that asks a LLM agent to...
300. Three-dimensional Conditional Diffusion Models for Cosmological 21 cm Lightcone Emulation ​
Author: Bin Xia, John H. Wise
Published: 9/2/2026, 4:00:00 AM
Categories: astro-ph.IM, astro-ph.CO, cs.LG
arXiv:2605.29016v2 Announce Type: replace-cross Abstract: We investigate conditional diffusion modeling for three-dimensional 21 cm lightcone emulation, focusing on cubes with a sky-plane size of $64\times64$ and a line-of-sight depth up to 1024 cells. Relative to earlier 2D studies, the 3D setting ...
301. HiMPO: Hindsight-Informed Memory Policy Optimization for Less-Entangled Credit in Long-Horizon Agents ​
Author: Jiangze Yan, Yi Shen, Wenjing Zhang, Jieyun Huang, Zhaoxiang Liu, Ning Wang, Kai Wang, Shiguo Lian
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2606.16285v2 Announce Type: replace-cross Abstract: Long-horizon agents rely on memory mechanisms to compress interaction history, but optimizing memory writing faces a distinct credit assignment challenge: a memory update may be rewarded or penalized due to downstream tool failures, noisy obs...
302. ReproRepo: Scaling Reproducibility Audits with GitHub Repository Issues ​
Author: Shanda Li, Qiuhong Anna Wei, Jingwu Tang, Valerie Chen, Nihar B Shah, Tim Dettmers, Yiming Yang, Ameet Talwalkar
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2606.18237v2 Announce Type: replace-cross Abstract: Reproducing research results from papers and released code is central to scientific progress. Existing works have introduced benchmarks to evaluate whether LLM agents can assist with reproducibility, but they are difficult to scale due to the...
303. Explicit Interaction Architectures for Dynamical Learning: A Controlled Study of Structural Inductive Bias ​
Author: Augusto Sarti
Published: 9/2/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2606.19101v2 Announce Type: replace-cross Abstract: We investigate a structure-first approach to dynamical learning in which the organization of stateful interactions is prescribed explicitly rather than left entirely to a generic recurrent parameterization. We introduce causal recurrent units...
304. Closing the Operational Gap in Semantic Caching ​
Author: Aditeya Baral, Radoslav Ralev, Iliya Sotirov Zhechev, Srijith Rajamohan, Jen Agarwal
Published: 9/2/2026, 4:00:00 AM
Categories: cs.IR, cs.CL, cs.LG
arXiv:2606.19719v4 Announce Type: replace-cross Abstract: Semantic caching cuts LLM inference costs by serving a cached response to semantically similar queries. Standard practice evaluates these systems using PR-AUC, a metric that only measures how well scores rank and ignores whether they are usab...
305. Steer, Don't Solve: Training Small Critic Models for Large Code Agents ​
Author: Shubham Gandhi, Yiqing Xie, Atharva Naik, Ruichen Zhu, Carolyn Rose
Published: 9/2/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2606.21811v2 Announce Type: replace-cross Abstract: Coding tasks are typically complicated and require multiple capabilities, ranging from high-level planning to low-level implementation. While coding agents are optimized for the joint capabilities, individual capabilities such as high-level p...
306. EchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in Echocardiography ​
Author: Darya Taratynova, Ahmed Aly, Numan Saeed, Mohammad Yaqub
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2606.28164v2 Announce Type: replace-cross Abstract: Echocardiography is the most widely used non-invasive cardiac imaging modality, providing essential information for cardiovascular diagnosis. Interpreting an echocardiogram requires synthesizing complementary evidence across multiple heart vi...
307. Can LLMs Imagine Moral Alternatives Beyond Binary Dilemmas? ​
Author: Jongchan Choi, Nari Yang, Sung Soo Park, Jaemin Cho, Han Seoyoung, Haerin Shin, Jun-Hyung Park
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2606.31213v2 Announce Type: replace-cross Abstract: As LLMs increasingly serve as moral advisors and agents, they must address conflicts between competing values. Yet prior work on moral dilemmas overlooks a central aspect of human moral cognition: imagining alternatives beyond the given optio...
308. How Does Alignment Tuning Shape Representations of Sycophancy and Related Cue-Induced Biases in LLMs? ​
Author: Prakhar Gupta, Terry Jingchen Zhang, Florent Draye, Bernhard Sch"olkopf, Zhijing Jin
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.18114v2 Announce Type: replace-cross Abstract: Modern LLMs are alarmingly susceptible to surprisingly simple immaterial changes of input prompts: a casual hint, an incorrectly labeled few-shot example, or a fake prior assistant turn often flips an originally correct answer. We study where...
309. A Classifier That Teaches Itself: Self-Improving, Frozen-gate Training (SIFT) for Dynamic Document Classification ​
Author: Bogdan Raduta, Horia Velicu, Alexandru Preda, Serban Chiricescu
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.18358v2 Announce Type: replace-cross Abstract: Document classification is a solved problem in the laboratory and an unsolved one in the enterprise. The blocker is rarely model architecture; it is the labeling project that must precede a model and the institutional fear of letting a model ...
310. Backspace as a Natural Experiment: An Accelerated Failure Time Model of Selective Post-Error Motor Impairment in Parkinsons Disease ​
Author: Navin Bondade
Published: 9/2/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.LG
arXiv:2607.24796v2 Announce Type: replace-cross Abstract: Parkinson's disease (PD) selectively impairs distinct stages of motor control. Using backspace events as natural error-correction episodes in the public neuroQWERTY MIT-CSXPD dataset (n=57 subjects, 27 PD with UPDRS-III scores), we test wheth...
311. Field-Aware Agent Skill Retrieval ​
Author: Paimon Goulart, Liang Wu, Kelly Wan, Evangelos E. Papalexakis, Liangjie Hong
Published: 9/2/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.02880v3 Announce Type: replace-cross Abstract: As lifelong learning agents accumulate lifelong growing skill banks, retrieving the correct skill becomes an increasingly important bottleneck. Most current skill retrieval methods treat each skill as one flat document by concatenating fields...
312. Coordinate-Residual Physics-Driven Neural Network for Inverse Scattering Imaging ​
Author: Yutong Du, Zicheng Liu, Bo Qi, Yali Zong, Peixian Han
Published: 9/2/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG, physics.app-ph
arXiv:2608.09382v2 Announce Type: replace-cross Abstract: Electromagnetic inverse scattering is a nonlinear and ill-posed computational imaging problem, where accurate reconstruction is challenging due to measurement limitations, noise, and high computational costs, especially for 3-D imaging. Altho...
313. Logarithmic-Free Moment and Generalization Bounds for Uniformly Stable Algorithms ​
Author: Thanh Nguyen-Cung, Binh T. Nguyen
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR, math.ST, stat.TH
arXiv:2608.09870v2 Announce Type: replace-cross Abstract: Uniform stability is a classical tool for controlling the generalization error of a learning algorithm. Bousquet, Klochkov, and Zhivotovskiy (2020) showed that the problem can be reduced to a moment inequality for a sum of weakly interacting ...
314. Debiased Inference for AI-Generated Data without Gold-Standard Labels: Identification via Multiple Imperfect Measurements ​
Author: Naoki Egami, Sooahn Shin
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ME, cs.AI, cs.CL, cs.LG, stat.ML
arXiv:2608.18294v2 Announce Type: replace-cross Abstract: An increasing number of scholars use AI to measure variables they subsequently include in downstream analyses. Although AI-measured variables are often analyzed as if observed without error, ignoring prediction errors in automated measurement...
315. ToSCA: Leveraging Hierarchical Reinforcement Learning on Temporal and Strategic Abstractions of Conversational Agents ​
Author: Xiaoyu Wang, Qingqing Gu, Yue Zhao, Teng Chen, Yuqi Cao, Xiaokai Chen, Hongyan Li, Luo Ji
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.HC, cs.LG
arXiv:2608.21969v3 Announce Type: replace-cross Abstract: Humans naturally exhibit multiple forms of abstraction in reasoning and interaction, including temporal abstraction across decision timescales and strategic abstraction over communicative intents. Inspired by these complementary abstractions,...
316. MRMAD: A Multi-Round Multi-Audio Benchmark for Evaluating Acoustic Degradation Perception in Large Audio-Language Models ​
Author: Yize Li, Ningyuan Yang, Sile Yin, Sindhuja Thogarrati, Sung-En Chang, Andrew C. Singer, Xue Lin, Chuan-Che Huang, Shuo Zhang
Published: 9/2/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS
arXiv:2608.22236v2 Announce Type: replace-cross Abstract: Large audio-language models (LALMs) have shown promising progress in understanding speech, music, and general sound events, yet their ability to reason about how audio signals are degraded remains underexplored. Existing benchmarks primarily ...
317. Common-Center Geometry and Certified Radial Reconstruction for Energy-Form Full Conformal Regions ​
Author: Yiheng Feng
Published: 9/2/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2608.24964v2 Announce Type: replace-cross Abstract: This note studies the geometry of full conformal prediction (FullCP) regions generated by an empirical energy-form pairwise score. Candidate-score convexity alone does not guarantee connected FullCP regions, even for empirical averages of los...
318. AutoScientist-Quant: Self-Evolving Coding Agents for Automatic Research in Quantitative Investment ​
Author: Zongqian Li, Yaoyiran Li, Yaohui Guo, Ming Zhang, Nigel Collier, Eugene Ie
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.28632v2 Announce Type: replace-cross Abstract: Large language model agents can discover alphas, yet current methods have three weaknesses. The search cannot adapt during the run, automation usually ends at alpha generation while library selection and model choice stay manual, and alpha di...
319. Hyper-Fold: Exploring the Expressive Limit of Sequence-Geometry Learning for Proteins via Hypergraph Modeling ​
Author: Yifan Feng, Guanjie Cheng, Shihui Ying, Shaoyi Du, Yue Gao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.29207v2 Announce Type: replace-cross Abstract: Protein structure modeling rests on a single computational primitive: the interaction between what a residue is (sequence content) and where it sits (three-dimensional geometry). What is the expressive limit of this layer class? We show that ...
320. Validating FKG.in: Soundness Assessment in LLM-Augmented Indian Food Knowledge ​
Author: Saransh Kumar Gupta, Armaan Shah, Lipika Dey, Partha Pratim Das, Ramesh Jain
Published: 9/2/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.IR, cs.LG
arXiv:2608.29249v2 Announce Type: replace-cross Abstract: The online culinary ecosystem is increasingly populated by recipe content generated, modified, or summarized by Large Language Models (LLMs). While often plausible, such outputs may contain hallucinated ingredients, misrepresented quantities,...
321. TACS: Trajectory-Aware Candidate Selection for LLM Jailbreak Suffix Optimization ​
Author: Shiliang Xiao
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.29564v2 Announce Type: replace-cross Abstract: Gradient-based jailbreak suffix optimization methods typically update the suffix by retaining the candidate with the lowest current loss. We show that this seemingly natural design is fundamentally myopic: candidates that look better under th...
322. BiG-SURE - Bipartite Graph for Semantic Uncertainty and Reliability Estimation of LLMs ​
Author: Debarpan Bhattacharya, Malay Phadke, Sriram Ganapathy
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.30646v2 Announce Type: replace-cross Abstract: Reliable uncertainty estimation is a crucial requirement for deploying large language models (LLMs) and vision-language models (VLMs) in safety-critical settings, especially when the model parameters are not accessible (black-box). We propose...
323. TopoCompress: Long Context Compression via Graph-Wired Semantic Trajectories ​
Author: Daniel Agyei Asante, Yang Li
Published: 9/2/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.30811v2 Announce Type: replace-cross Abstract: Long-context compression is essential for reducing the cost and latency of large language model inference. However, existing methods can fragment important evidence, require additional training or alignment, and often depend on the target mod...