Skip to content

arXiv cs.LG - 2026-09-03 ​

278 items collected.


1. WMLLM: Self-Evolving Optimization Agents via Predict-Then-Act World Modeling ​

Author: Zhongzheng Li, Qingsong Ran, Shikun Feng, Nian Ran, Wenhao Li, Xiaoyuan Zhang, Yue Wang, Xiaoguang Zhao
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.01608v1 Announce Type: new Abstract: Black-box optimization problems remain challenging because of large, weakly structured, and high-dimensional search spaces. Existing methods often suffer from poor sample efficiency because they rely on direct candidate generation or trial-and-error re...

📖 Read original article


2. DiDrive: A Risk-Aware Hierarchical Diffusion Framework for Safe Offline Reinforcement Learning in Autonomous Driving ​

Author: Qisong Guo, Jingtang Chen, Zhilin Chen, Pei Xu, Mingjian Fu, Wenxi Liu, Yuanlong Yu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2609.01609v1 Announce Type: new Abstract: While diffusion models effectively capture multimodal behavioral priors for autonomous driving, offline reinforcement learning (RL) policies remain susceptible to distribution shift, heavy-tailed risk signals, out-of-distribution (OOD) action generatio...

📖 Read original article


3. Prompt-Space Meta-Learning Does Not Transfer Across Users: A Frozen-LLM Negative Result ​

Author: Liam Byrne, David Dylan, Orla Fitzgerald, Eoin Doyle, Ciara Nolan, Padraig Lynch, Sinead Gallagher
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01615v1 Announce Type: new Abstract: Personalizing a frozen large language model (LLM) to individual users is often framed as a meta-learning problem in prompt space: each user is a task, and one seeks a shared natural-language adaptation policy that, given a handful of the user's labeled...

📖 Read original article


4. Efficient Context-Limited Telescope Bibliography Classification for the WASP-2025 Shared Task Using SciBERT ​

Author: Madhusudhana Naidu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.IM

arXiv:2609.01647v1 Announce Type: new Abstract: The creation of telescope bibliographies is a crucial part of assessing the scientific impact of observatories and ensuring reproducibility in astronomy. This task involves identifying, categorizing, and linking scientific publications that reference o...

📖 Read original article


5. CliffRank: A Dual-Branch Framework for Activity-Cliff Ranking Prediction ​

Author: Kewei Li, Rongying Zhang, Peiyu Yang, Zhongjian Wang, Qiuchen Zhao, Lan Huang, Fengfeng Zhou
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.BM

arXiv:2609.01673v1 Announce Type: new Abstract: Activity-cliff ranking remains difficult because local structural changes can cause large activity differences, while high-quality data that resolve the underlying mechanisms remain limited. To use available activity labels more effectively, we combine...

📖 Read original article


6. Sim2Signal: Sim-to-Real Benchmarks for Traffic Signal Control ​

Author: Ferdous Al Rafi, Susrik Mukherjee, Latika Liladhar Dekate, Jennifer Yawa Lavoe, Huaiyuan Yao, Shlok Mohanty, Longchao Da, Xuesong Zhou, Hua Wei
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01676v1 Announce Type: new Abstract: Reinforcement learning achieves strong traffic signal control performance in simulation, yet policies trained in simulators often fail once deployed in the real world, a failure known as the Sim-to-Real gap. When RL is applied to traffic signal control...

📖 Read original article


7. A Survey on Self-Improving Test-Time Intelligence: Feedback-Driven Adapting, Learning, and Scaling at Inference ​

Author: Shuaicheng Niu, Guohao Chen, Yaofo Chen, Zhiquan Wen, Jinwu Hu, Zeshuai Deng, Deyu Chen, Shuhai Zhang, Renjie Chen, Zihao Lian, Shoukai Xu, Gang Dai, Yunbei Zhang, Wei Luo, Yifan Zhang, Mingkui Tan, Cheng Deng
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01679v1 Announce Type: new Abstract: The ability of AI systems to improve their behavior during deployment is becoming increasingly important. As inference moves beyond the static execution of a fixed trained model, a growing body of work studies how models can refine their behavior on th...

📖 Read original article


8. Reinforcement Learning and Rule-Based Peer-to-Peer Pricing in Residential PV-BES Communities ​

Author: Pablo Benalcazar, Maciej Kalka, Wilian Guam'an, Jacek Kami'nski
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2609.01680v1 Announce Type: new Abstract: This paper compares rule-based and learning-based pricing mechanisms for peer-to-peer (P2P) electricity trading in residential photovoltaic communities. The rule-based benchmarks comprise bill-sharing as an ex post allocation mechanism, the mid-market ...

📖 Read original article


9. Median-of-Means as an Extremal Convex Estimator and a Nonconvex Route to the Trimmed Oracle ​

Author: Angshul Majumdar
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01689v1 Announce Type: new Abstract: We revisit median-of-means estimation from a deterministic optimization viewpoint and develop a family of block-Lp estimators for robust learning with heavy-tailed and adversarially corrupted data. In a block contamination model with at least a fractio...

📖 Read original article


10. Tri-Band Channel Measurement-Enabled Multi-Layer Digital Twin for Terahertz Wireless Data Centers ​

Author: Mingjie Zhu, Ziming Yu, Guangjian Wang, Chong Han
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT

arXiv:2609.01699v1 Announce Type: new Abstract: The rapid growth of AI computing has driven increasing demands for flexible and high-capacity data-center interconnections. Owing to its ultra-wide bandwidth and high spatial reuse capability, terahertz (THz) communication has emerged as a promising so...

📖 Read original article


11. Generative Diffusion Surrogates with Analytical Variance Schedule ​

Author: Patrick Reichherzer, Gianluca Gregori, David N. Hosking, Subir Sarkar
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.IM, physics.plasm-ph

arXiv:2609.01705v1 Announce Type: new Abstract: Stochastic transport describes physical systems in which an initially structured distribution spreads under unresolved forcing, scattering, or heterogeneous media. Useful surrogates for such systems should be probabilistic, time-resolved, and able to r...

📖 Read original article


12. RecKAN: Kolmogorov-Arnold Networks with a Learnable Recursive Polynomial Basis ​

Author: Amirhosein Azarpour
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.01729v1 Announce Type: new Abstract: Kolmogorov--Arnold Networks (KANs) replace the fixed scalar weights of a standard network with learnable univariate functions on each edge, but existing variants still fix the \emph{basis} that those functions are built from: B-splines, Chebyshev polyn...

📖 Read original article


13. CAT-Flow: Curvature-Adaptive sTeps for Flow Matching ​

Author: Qinchan Li, Pedro Cisneros-Velarde, Keru Fu, Samuel Antunes Miranda, Sharan Vaswani, Hao Zhang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01746v1 Announce Type: new Abstract: Flow Matching has emerged as a leading framework for generative modeling, powering state-of-the-art systems such as FLUX and Stable Diffusion 3.5. However, the iterative nature of its ODE-based sampling process creates a fundamental efficiency bottlene...

📖 Read original article


14. A Study of Conditional Diffusion Models for Open-Loop Control under Dry Friction and Stiction ​

Author: Eric Aislan Antonelo
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01756v1 Announce Type: new Abstract: Diffusion models have recently emerged as expressive generative priors for planning and control. This paper studies Action Diffusion, an action-sequence diffusion formulation used as an open-loop proposal distribution for a point-mass system with dry f...

📖 Read original article


15. Toward Explainable and Policy-Aware AI for Carbon Credit Price Prediction: A Research Framework for Emerging Carbon Markets ​

Author: Summaiya Unnisa Begum, Mohammed Nadeem Ullah, Mohammed Abdul Ghani Khan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01765v1 Announce Type: new Abstract: Carbon markets put a price on emissions, yet that price remains hard to forecast. Work in this area clusters on the EU and Chinese schemes, compresses regulatory text into a sentiment score, and reports accuracy without calibration or explanation stabi...

📖 Read original article


16. Emergence of Fibrations, Compression, and Symmetry Breaking in Artificial Neural Networks ​

Author: Osvaldo M Velarde, Lucas C Parra, Alireza Hashemi, Hernan A Makse
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01768v1 Announce Type: new Abstract: Artificial neural networks are often regarded as powerful yet opaque black boxes. Here, we demonstrate that learning in deep neural networks generates local symmetries known in graph theory as fibrations and coverings. We prove that covering symmetries...

📖 Read original article


17. D-FROST: Decentralized Federated pRompt-tuning via Optimal tranSporT for Non-IID and Imbalanced Data ​

Author: Quan Minh Nguyen, Hoang M. Ngo, Trong Nghia Hoang, My T. Thai
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01802v1 Announce Type: new Abstract: Prompt tuning provides a parameter-efficient way to adapt foundation models (FMs) by freezing the pretrained backbone and updating only a small set of learnable prompts. This property makes prompt tuning especially suitable for decentralized federated ...

📖 Read original article


18. hLLM: Single Pass Decoding for Generative Reranking ​

Author: Emil Laftchiev, Prachi Agrawal, Moe Kayali, Bixing Yan, Qi Xu, Zijie Lei, Chen Qiu, Zhi Hua, Ke Li, Luke Simon
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR

arXiv:2609.01807v1 Announce Type: new Abstract: Large language models (LLMs) achieve state-of-the-art generative ranking quality, but the ranking they produce must be decoded, and autoregressive decoding spends one sequential forward pass per emitted token. We observe that the only tokens a ranker m...

📖 Read original article


19. Import What You Need: Learning When and How to Augment EHR Graphs with External Knowledge ​

Author: Chen Chen, Mohsen Nayebi Kerdabadi, Dongjie Wang, Mei Liu, Zijun Yao
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.01839v1 Announce Type: new Abstract: Longitudinal prediction from electronic health records (EHRs) is limited by the sparsity and irregularity in patient trajectories, and knowledge augmentation with external knowledge graphs (KGs) offers a promising way to alleviate these issues. However...

📖 Read original article


20. OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation ​

Author: Yunqin Zhu, Feng Qiu, Yao Xie
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.01896v1 Announce Type: new Abstract: Power-outage planning requires scenarios before an event occurs. These scenarios must represent uncertainty in magnitude, timing, and duration while preserving temporal dependence. However, severe events are rare, and data from any single region contai...

📖 Read original article


21. CRISP: Cliff-awaRe Input-adaptive Sparse Prefilling with Structural-Mass-Motivated Routing ​

Author: Huu Huy Nguyen, Chien Van Nguyen, Franck Dernoncourt, Ryan A. Rossi, Linh Ngo Van, Jieyang Chen, Thien Huu Nguyen
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2609.01925v1 Announce Type: new Abstract: The attention prefilling phase of long-context LLM inference scales quadratically, making self-attention a severe computational bottleneck. Traditional sparse attention methods mitigate this through fixed patterns or offline profiling, but lack the fle...

📖 Read original article


22. OR-Transformer: Scaling Real-Time Decision-Making to 1,000 Items ​

Author: Shuze Daniel Liu, David Simchi-Levi, Claire Chen, Chutong Gao, Shangtong Zhang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01933v1 Announce Type: new Abstract: Modern supply chain operations can require coordinating replenishment across thousands of heterogeneous items under correlated stochastic demand, heterogeneous lead times, and shared fixed ordering costs, yielding observation spaces exceeding $10^4$ di...

📖 Read original article


23. Refining Heuristic-Based Bitcoin Address Clustering with Graph Neural Networks ​

Author: Hugo Schnoering, Roman Bresson, Michalis Vazirgiannis
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01942v1 Announce Type: new Abstract: Bitcoin's pseudonymous nature makes it challenging to analyze user-level activity, since a single user may control multiple identifiers (addresses). Existing heuristic-based methods attempt to identify addresses belonging to the same user, but they oft...

📖 Read original article


24. On-Policy Distillation Meets Off-Policy GRPO: Training Compact Instruction-Following Rerankers ​

Author: Vignesh Prabhakar, Jialing Pan, Anil Babu Ankisettipalli
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.01947v1 Announce Type: new Abstract: Compact instruction-following rerankers are attractive for deployment, but conventional distillation pipelines typically train students by offline imitation of teacher outputs on a fixed set of examples, constraining supervision to the teacher's observ...

📖 Read original article


25. Convergence Theory of Knowledge Distillation in Asynchronous P2P Gossip Learning Network ​

Author: Lucas Qingyang Fang, Tiyao Liu, Jinhao Jing, Zeji Li, Kaijie Chen, Harikrishna Kuttivelil, Katia Obraczka
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.01952v1 Announce Type: new Abstract: Decentralized, serverless learning increasingly connects devices running different architectures, where the standard tool, decentralized SGD, is undefined as models with different parameter counts cannot be averaged. Knowledge distillation (KD) exchang...

📖 Read original article


26. InKAN: B-Spline KANs via Truncated Power Form ​

Author: Naveen Mysore
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2609.01956v2 Announce Type: new Abstract: Kolmogorov-Arnold Networks (KANs) place learnable B-spline activations on network edges rather than fixed activations on nodes. The standard Cox-de Boor recursion evaluates these activations through $k$ sequential passes for degree-$k$ splines, consumi...

📖 Read original article


27. A Unified Particle Filter LSTM for Data-Driven Process Simulation ​

Author: Parvin Malekzadeh, Opher Baron, Dmitry Krass
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2609.01967v1 Announce Type: new Abstract: Data-driven process simulation aims to generate realistic case trajectories from historical event logs without requiring an explicitly specified model of the underlying dynamics. Deep sequence models can capture complex temporal dependencies through ne...

📖 Read original article


28. CAHR-Net: Condition-Adaptive Hysteresis Reconstruction for Compact and Interpretable Magnetic Core Loss Modeling ​

Author: Chunye Gong, Cong Yao
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01991v1 Announce Type: new Abstract: Magnetic core loss originates in the hysteresis loop: the energy dissipated per excitation cycle equals the loop area, and frequency, temperature, and waveform shape set the loss by reshaping the loop geometry. Most existing models let these conditions...

📖 Read original article


29. Train What You Deploy: Closing the MLP Reachability Gap in Low-Rank Clone Distillation ​

Author: Wenhui Chen, Zhifeng Li, Jie Zhou, Navan Preet Singh, Madalina Ciobanu, Chenghua Wang, Qingqing Mao, Ritankar Das
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2609.02006v1 Announce Type: new Abstract: A compressed student has two shapes that need not agree: the weight it deploys at inference and the weight family its training can reach. We show that a state-of-the-art weight-inheritance distiller, Low-Rank Clone (LRC), deploys a full-width student M...

📖 Read original article


30. Source-Free Class Relearning: Diagnosing Forgetting in Class Unlearning ​

Author: Zahra Dehghani, Pablo Piantanida, Mohammadhadi Shateri
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2609.02018v1 Announce Type: new Abstract: Class unlearning aims to remove a model's ability to recognize designated forget classes while preserving performance on retain classes. However, low forget accuracy after unlearning does not necessarily mean the class structure has been erased. Approx...

📖 Read original article


31. Act More, Decide Less: Skill-Guided Adaptive Action Chunking for Long-Horizon LLM Agents ​

Author: Yanting Yang, Can Jin, Jinman Zhao, Jiahao Wu, Yang Zhou, Zhepeng Wang, Zhendong Wang, Mu Zhou, Dimitris N. Metaxas
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02042v1 Announce Type: new Abstract: Large language model (LLM) agents for long-horizon interactive tasks typically follow a ReAct-style protocol, issuing one primitive action per LLM round. While this enables frequent replanning, it is inefficient for long-horizon tasks where many rounds...

📖 Read original article


32. The Dynamics of Continuous Mixture Collapse in Language Models ​

Author: Ali Backour
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2609.02049v1 Announce Type: new Abstract: LLMs latent-state reasoning methods replace discrete intermediate tokens with continuous states, such as weighted mixtures of token embeddings, to retain multiple possible reasoning directions rather than committing to one. Yet pretrained language mode...

📖 Read original article


33. DynG-Diff: A State-Aware Dynamic Guidance Diffusion Framework for Probabilistic Time Series Forecasting ​

Author: Zhente Zhang, Zhengwei Ni, Wei Fan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02068v1 Announce Type: new Abstract: Probabilistic multivariate time series (MTS) forecasting is crucial for modeling complex dynamical systems. However, existing diffusion-based methods rely on task-specific conditional paradigms that lack flexibility and struggle with inherent "informat...

📖 Read original article


34. XMerge: Cross-Axis Selection and Reconstructive Layer Merging for LLM Depth Compression ​

Author: Jundong Hu, Shekar Ramachandran
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2609.02083v1 Announce Type: new Abstract: Removing complete transformer layers preserves a standard serving architecture, but existing depth-compression methods can lose substantial quality, and the loss varies unpredictably across models. We introduce XMerge, a post-training method with two c...

📖 Read original article


35. TC-Next: Zero-Shot Multimodal Cyclone Forecasting ​

Author: Zhe Wang, Sijie Chen, Yiming Luo, Daehyun Kim, Chien-Yi Chang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2609.02085v1 Announce Type: new Abstract: We present TropicalCycloneNext (TC-Next), a multimodal deep learning model that forecasts tropical cyclone track and intensity at $6$-$24$ h leads by leveraging a foundation model's forecast fields of atmospheric kinematic and thermodynamic fields and ...

📖 Read original article


36. Compositional Spectral Prompts for LLM-based Online Time Series Forecasting ​

Author: Seungyoon Choi, Hyunchul Kim, Jae-Gil Lee, Chanyoung Park
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02093v1 Announce Type: new Abstract: To address the sequential and evolving nature of time series, the Online Time Series Forecasting (OTSF) task has been extensively studied in multiple domains. Existing research focuses on adapting to non-stationary environments by employing memory buff...

📖 Read original article


37. Federated LoRA Adaptation of BiomedCLIP Across Four International Chest X-Ray Cohorts ​

Author: Sanjaya Poudel, Nirajan Kunwor, Manish Dhakal, Debesh Jha, Sunil Kumar Gaire
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2609.02101v1 Announce Type: new Abstract: Federated learning (FL) lets institutions train a shared model without exchanging data, and Low-Rank Adaptation (LoRA) makes this practical at scale by communicating only compact low-rank updates. Biomedical imaging is a compelling setting for this com...

📖 Read original article


38. A Unified Rate-Distortion Perspective on Vector, Product, and Scalar Quantization ​

Author: Xianghong Fang, Wenlong Mou, Yuan Yuan, Dehan Kong, Tim G. J. Rudner
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2609.02107v1 Announce Type: new Abstract: Discrete visual tokenization, predominantly driven by vector, scalar, and product quantization, lacks a unified conceptual framework for understanding quantization tradeoffs. In this paper, we propose a unified rate--distortion perspective on modern di...

📖 Read original article


39. A Computational Comparison of Fourier Spectral Differentiation and Spatial Automatic Differentiation in Periodic Physics-Informed Neural Networks ​

Author: Xilai Liang, Zhao Zhang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02110v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) commonly evaluate the spatial derivatives appearing in partial differential equation residuals using automatic differentiation (AD), whose computational and memory costs can become substantial when multiple or h...

📖 Read original article


40. Scalable Bayesian Optimization of Composite Functions for Image-Based Inverse Problems in Materials Characterization ​

Author: Dasol Yoon, Poompol Buathong, Chia-Hao Lee, Yujia Zhang, David A. Muller, Peter I. Frazier
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci

arXiv:2609.02126v1 Announce Type: new Abstract: Estimating physical parameters from scientific images is a common inverse problem in materials characterization that often relies on expensive physics-based simulations. In electron microscopy, specimen thickness and crystal mistilt are critical parame...

📖 Read original article


41. Online Non-Monotone DR-Submodular Maximization Matching the Offline $0.401$ Factor ​

Author: Vaneet Aggarwal, Yiyang Lu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CC, stat.ML

arXiv:2609.02145v1 Announce Type: new Abstract: We study online maximization of nonnegative, non-monotone DR-submodular functions over compact convex down-closed subsets of the $d$-dimensional unit cube. The best known constructive offline approximation factor is $0.401$ under the corresponding meta...

📖 Read original article


42. Exact Limits of Random Projections for Preserving Geometry: Distance Recovery, Nearest-Neighbor Rankings, and Covariance Shape in Gaussian Models ​

Author: Piyush Sao
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, cs.NA, math.IT, math.NA, math.PR, math.ST, stat.TH

arXiv:2609.02155v1 Announce Type: new Abstract: The Johnson-Lindenstrauss (JL) lemma guarantees that a random projection of $n$ points to $m=O(\varepsilon^{-2}\log n)$ dimensions preserves pairwise squared distances within relative error $\varepsilon$ with high probability, and this dimension order ...

📖 Read original article


43. GeoSPRINT: Geometric Redundancy-Aware Step Pruning for Inference in Diffusion Trajectories ​

Author: Arpita Joshi
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02160v1 Announce Type: new Abstract: Diffusion models achieve high sample quality but remain expensive at inference time because sampling requires many sequential neural function evaluations (NFEs). Existing acceleration methods either use fixed step-skipping schedules, adapt step sizes b...

📖 Read original article


44. DMRL: Document-Mediated Reinforcement Learning for Skill Optimization in Advertising Recommendation ​

Author: Wei Zhang, Hongji Li, Song Sun, Peng Yu, Xue Yang, Lei Zhao, Peng Jiang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02170v1 Announce Type: new Abstract: Advertising recommendation requires continuously tuning complex system parameters while balancing commercial returns and user experience. Recent work has introduced large language models (LLMs) with skill documents to assist this labor-intensive proces...

📖 Read original article


45. Learning the Constitutive Behavior of Materials via Neural Operators and Causal Attention: Case Studies in Plasticity and Damage ​

Author: Rishabh Arora, Lisa Scheunemann, Tim Brepols, Shahed Rezaei
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CE

arXiv:2609.02194v1 Announce Type: new Abstract: Classical constitutive modeling of path-dependent inelastic materials relies on internal state variables whose evolution equations must be postulated based on domain knowledge and calibrated against experimental data. However, in many practical setting...

📖 Read original article


46. SMart: A Multi-source Multi-phase Time Series Representation Transfer Framework ​

Author: Fang He, Wang-chien Lee
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02203v1 Announce Type: new Abstract: Time series representation learning (TSRL) has attracted growing research interests in recent years. Two recent explorations in TSRL are: i) exploiting a transformer-based framework to learn time series; ii) instead of using only the targeted dataset, ...

📖 Read original article


47. Recursive Value Learning for Long-Horizon Offline Goal-Conditioned RL ​

Author: Hyeonseong Jeon, Youngwoon Lee
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2609.02237v1 Announce Type: new Abstract: Scaling offline goal-conditioned reinforcement learning (GCRL) to long-horizon tasks is difficult because (1) long-range value learning depends on shorter-range estimates that may still be inaccurate, and (2) max-based value backups can amplify overest...

📖 Read original article


48. Similarity-Aware Personalized Federated Learning in Heterogeneous Environments ​

Author: Arun Kumar A V, Sunil Gupta, Dang Ngyuen, Bao Duong, Dat Phan Trong
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02241v1 Announce Type: new Abstract: Federated Learning (FL) allows decentralized clients to train models collaboratively while preserving data privacy. However, distribution mismatch across clients often leads to poor global generalization and degraded local client-level performance. In ...

📖 Read original article


49. CAPTURE: Disentangling Preference Drift from Memory Poisoning in Personalized LLM Agents ​

Author: S M Asif Hossain, Ruksat Khan Shayoni, Md Kishor Morol
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02265v1 Announce Type: new Abstract: Personalized language agents use persistent memory to adapt to users over time, but the same mechanism creates an attack surface. When new information conflicts with stored preferences, an agent must distinguish genuine preference drift from temporary ...

📖 Read original article


50. Entangled Representations Amplify Collateral Damage in Unlearning ​

Author: Ev\v{z}en Wybitul, Tim G. J. Rudner, Christian Schroeder de Witt
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2609.02285v1 Announce Type: new Abstract: A long-held intuition in interpretability research is that representational entanglement, the sharing of structure between knowledge domains in a neural network, makes unlearning harder. While the intuition is widespread, it has never been directly tes...

📖 Read original article


51. SEAL: Reinforcing Global Safety in Mixture-of-Experts through Shared Expert ALignment ​

Author: Qingyu Meng, Yiwei Zha, Jiahuan Pei, Koen Hindriks, Herbert Bos, Min Chen
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2609.02293v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) is a scaling architecture for large language models that activates only a small subset of expert modules per token, enabling massive parameter growth with nearly constant computation. Recent Hybrid MoE architecture adds \textit...

📖 Read original article


52. Bayes-Optimal BER and AUC: Estimation and Evaluation of Estimators ​

Author: Ryota Ushio, Takashi Ishida, Masashi Sugiyama
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2609.02304v1 Announce Type: new Abstract: A fundamental quantity in machine learning is the optimal performance achievable by any model on a given task. Estimating this quantity allows us to distinguish the irreducible part of the error from a deficiency of the model, telling us how much room ...

📖 Read original article


53. What Is Worth Representing? Representational Empowerment for Continual Model Construction ​

Author: Fei Dai, Hanqi Zhou, Alison Gopnik, Charley Wu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02322v1 Announce Type: new Abstract: The first problem of modeling the world is not just estimating the right parameters or causal structure, but deciding what should be represented at all. We frame this problem as continual model construction: an agent maintains an environment-specific m...

📖 Read original article


54. AGI Maze Prediction Datasets: A Compact Benchmark for Learning World Dynamics with Transformers ​

Author: Alexey Potapov
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02339v1 Announce Type: new Abstract: World modeling requires a predictive model to maintain and update an internal state adequate for reasoning about the consequences of actions. We introduce the AGI Maze Prediction Datasets and Benchmark, a lightweight controlled testbed for studying thi...

📖 Read original article


55. Percolation Dynamics in Optimization : Variance Cascades and Discrete Scale Invariance ​

Author: Sai Niranjan Ramachandran, Suvrit Sra
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, cond-mat.stat-mech, cs.AI

arXiv:2609.02373v1 Announce Type: new Abstract: We study the dynamics of Stochastic Gradient Descent (SGD), which is known to steer deep neural networks toward invariant sets that correspond to simpler subnetworks. How this steering unfolds over time remains poorly understood. We answer this by mode...

📖 Read original article


56. Evidence for Shared Routing Geometry and Dynamics in Sparse Mixture-of-Experts ​

Author: Kirill Labzin, Stepan Kulibaba, Artem Dzhalilov, Artem Gorokhov
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02404v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) models use an independently parameterized router at each sparse layer to select experts for every token. Prior work has shown that routing decisions across depth can often be predicted from earlier routing signals, sugge...

📖 Read original article


57. Coverage, Not Targeting: A Structural Regime in Multi-Turn Agent Credit Assignment ​

Author: Chenyu Zhou, Qiliang Jiang, Shuning Wu, Xu Zhou
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02417v1 Announce Type: new Abstract: Multi-turn agentic RL increasingly treats credit assignment as a targeting problem: given a terminal verifiable reward, per-turn methods localize credit onto the turns that mattered. We identify the structural quantity that predicts when this is the ri...

📖 Read original article


58. IFW-BLS: Dual-Robust Broad Learning System with Intuitionistic Fuzzy Wave Loss ​

Author: Mushir Akhtar, M. Tanveer
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02422v1 Announce Type: new Abstract: Broad Learning System is an efficient randomized learning model that expands network width through feature and enhancement nodes and estimates the output weights without deep backpropagation. Its standard least-squares training, however, is vulnerable ...

📖 Read original article


59. Towards One-for-All Robustness Across a Continuum of Threat Levels ​

Author: Zhichao Hou, Xiaorui Liu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02440v1 Announce Type: new Abstract: Adversarially robust models often overfit to a specific attack budget, necessitating multiple specialized models for diverse and dynamic adversarial environments, a strategy that becomes fundamentally intractable as the threat space grows. This raises ...

📖 Read original article


60. CACTUS: Mask-Guided Semantic Clean-Label Backdoors in Decentralized Federated Learning ​

Author: Chao Feng, Burkhard Stiller
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2609.02450v1 Announce Type: new Abstract: Semantic triggers in federated learning (FL) can be less conspicuous than synthetic patches, but sample-dependent placement may weaken backdoor implantation across aggregation rounds. This challenge is compounded in decentralized FL (DFL), where topolo...

📖 Read original article


61. Scalable Kronecker-Fisher Approximation: Efficient Hessian Analysis for Billion-Parameter Language Models Compression ​

Author: Viacheslav Yusupov, Daria Cherniuk, Evgeny Frolov
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2609.02451v1 Announce Type: new Abstract: In this paper, we propose a scalable Kronecker-based approximation that captures cross-layer interactions without storing the entire Fisher matrix, enabling practical Hessian analysis for billion-parameter networks where full computation is infeasible....

📖 Read original article


62. DeepAffinity: Long-Term Aspect Preference Prediction in eCommerce using Small Language Models ​

Author: Yotam Eshel, Guy Hadad, Guy Feigenblat, Yuri M. Brovman, Matt Gearhart, Bracha Shapira
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02468v1 Announce Type: new Abstract: We explore predicting eCommerce user preferences for product aspects such as brand, size, and color - a task we define as Aspect Affinity. Solving this task improves customer understanding and enables fine-grained personalization in recommendation, sea...

📖 Read original article


63. RINSE: Robust Target-Time Normality Estimation for Zero-Shot Graph Anomaly Detection ​

Author: Taufikur Rahman Fuad, Md Abrar Jahin, Amir Hussain
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02497v1 Announce Type: new Abstract: Zero-shot graph anomaly detection seeks to deploy a detector trained on source graphs to unseen, unlabeled targets, yet domain shift can make source-derived notions of normality unreliable. We introduce RINSE (Robust Iterative Normality Self-Estimation...

📖 Read original article


64. Rethinking the Teacher-Student Framework for Test-Time Adaptation ​

Author: Damian S'ojka, Marc Masana, Bart{\l}omiej Twardowski, Sebastian Cygert
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02507v1 Announce Type: new Abstract: Test-Time Adaptation (TTA) has recently emerged as a promising strategy that allows the adaptation of pre-trained models to changing data distributions at deployment time, without access to any labels. To mitigate error accumulation, researchers have w...

📖 Read original article


65. Spectral Initialization and Scheduled Graph Smoothness for Uncertain Knowledge Graph Completion ​

Author: Md Abrar Jahin, Taufikur Rahman Fuad, Jay Pujara, Craig A. Knoblock
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02519v1 Announce Type: new Abstract: Uncertain knowledge graphs (UKGs) extend knowledge graphs by assigning each triple a continuous confidence score. Since most possible triples lack observed confidences, recent methods rely on semi-supervised learning to generate pseudo-labels. These me...

📖 Read original article


66. A Comparative Study of Graph Representations for GNN-Based Power Grid Control in L2RPN ​

Author: Adrian Degenkolb, Qiong Huang, Benjamin Sch"afer
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02538v1 Announce Type: new Abstract: Graph construction is a critical but underexamined design choice in deep reinforcement learning for power grid control. We present a controlled experimental comparison of different graph representations, including physical topology, electrical-sensitiv...

📖 Read original article


67. TrajMind: Chaining Role-Specialized LoRAs for Fast-and-Slow Collective Trajectory Anomaly Diagnosis ​

Author: Jiahao Wu, Zhenqun Yang, Chen Jason Zhang, Qing Li
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02540v1 Announce Type: new Abstract: Diagnosing collective anomalies from urban trajectories is increasingly important for traffic governance, as it reveals what happened, who was involved, and where and when the event occurred. Existing detectors efficiently produce scores or labels, whe...

📖 Read original article


68. Learn from Whoever Is Right: Answer-Verified Multi-Teacher Distillation for Multi-Domain LLMs ​

Author: Xixiang He, Xingming Li, Baiqi Wu, Qiyao Sun, Xuanyu Ji, Ao Cheng, Qingyong Hu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02548v1 Announce Type: new Abstract: Modern large language models (LLMs) rely on reinforcement learning to build strong capabilities in individual domains, but integrating those capabilities into a single deployable model remains challenging. By routing each sample to the teacher whose do...

📖 Read original article


69. ProbeMatchDTI: Probe-Driven Multi-Scale Biochemical Pattern Matching for Drug-Target Interaction Prediction ​

Author: Quan Hao, Mengyue Fan, Zifan Dong, Youru Li, Jianduo Zhao, Lechuan Xu, Hao Zhang, Fei Xia, Jigang Wang, Chong Qiu, Liguo Zhang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.02549v1 Announce Type: new Abstract: Drug-target interaction (DTI) prediction is an important task in AI-driven drug discovery. Although recent biochemical representation learning methods have improved DTI prediction, their passive feature aggregation tends to favor dominant molecular pat...

📖 Read original article


70. Online Reinforcement Learning in the Met Office Unified Model through Distributed Model-Agent Coupling ​

Author: Pritthijit Nath, Sebastian Schemm, Peter Haynes, Emily Shuckburgh, Mark Webb
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02566v1 Announce Type: new Abstract: Machine-learnt corrections can complement numerical weather prediction only if they adapt to the evolving model state while preserving dynamical consistency and numerical stability. To test this within a global forecasting model, we couple the Met Offi...

📖 Read original article


71. Source Distribution Estimation by Posterior Averaging ​

Author: Trung-Dung Hoang, Lisa M. Koch
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02622v1 Announce Type: new Abstract: Simulation-based science often requires a distribution over simulator parameters whose push-forward reproduces a set of real observations: this is the source distribution estimation (SDE) problem. Existing methods fit the source against a likelihood su...

📖 Read original article


Author: Guillaume M'erou'e, Fabien Gandon, Pierre Monnin
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02638v1 Announce Type: new Abstract: Knowledge graphs have become an important source of structured knowledge for Web applications, including search, question answering, and recommender systems. In these applications, link prediction can serve either as a prediction task itself or as a me...

📖 Read original article


73. Differentiable Electricity-Market Clearing for Gradient-Based Planning ​

Author: Luca Mungo, Maarten P. Scholl, Arnau Quera-Bofarull
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2609.02646v1 Announce Type: new Abstract: Planning a large data center is difficult because a facility big enough to matter changes the electricity prices it will pay. Those prices are set by market clearing, a constrained optimization problem solved anew in every operating condition. However,...

📖 Read original article


74. Unfolding the Leech Lattice: Fused Multi-Shell Decoding and VRAM Layouts for 2-Bit LLM Weights ​

Author: Pier-Jean Malandrino (Scub)
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02652v1 Announce Type: new Abstract: Leech-lattice vector quantization holds the strongest reported 2-bit quality under its own evaluation protocol. Its kernel decodes one shell; we found no implementation of the multi-shell decoder the rate requires. This paper supplies one and measures ...

📖 Read original article


75. H3DNAS: Hardware-Aware ONNX-Native 3D Point Cloud Model Compression ​

Author: Anchit Mulye, Rhythm Baghel, Sujay Kumar Ingle, Hardik Jain
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.NE

arXiv:2609.02684v1 Announce Type: new Abstract: Deploying 3D point cloud models on edge hardware such as the NVIDIA Jetson Orin Nano is severely constrained by compute and memory budgets. Existing compression methods require access to the model's original source code, rendering them inapplicable to ...

📖 Read original article


76. LoRA-TSD: Tangent-Space Spectral Descent for LoRA via Muon-Style Updates ​

Author: Dmitrii Andriianov, Andrey Veprikov, Aleksandr Beznosikov
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02734v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) is the standard way to fine-tune large models, yet when its two factors are trained independently, the update ignores the geometry of the low-rank weight change it induces. We introduce LoRA-TSD, an optimizer that treats ever...

📖 Read original article


77. Do Tabular Foundation Models Know Physics? Contamination, Units, and the Deterministic Limit ​

Author: Wassim Tenachi, Yashar Hezaveh, Laurence Perreault Levasseur, Pierre-Luc Bacon
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.IM

arXiv:2609.02766v1 Announce Type: new Abstract: Tabular foundation models (TFMs) learn to fill in tables the way language models fill in text, and tables are arguably the format in which most physical measurement arrives. Did they learn any physics in the process? They are Bayesian by construction, ...

📖 Read original article


78. Cliff: Learning Process Rewards from the First Mistake ​

Author: Peixuan Han, Runhui Wang, Ketan Ramaneti, Jie Hao, Gerald Friedland, Chris Kong
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02817v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for large language model (LLM) post-training, but its reliance on coarse outcome rewards leads to limited guidance on intermediate reasoning processes. Existing ap...

📖 Read original article


79. UE5M3 FP4 Block Scaling for Stable Language Model Pretraining ​

Author: Robert Hu, Carlo Luschi, Paul Balanca
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02846v1 Announce Type: new Abstract: Stable 4-bit floating-point (FP4) pretraining is difficult because the E2M1 payload represents only a narrow range of magnitudes. NVIDIA's Transformer Engine \nv{} recipe addresses this with current-tensor scaling, a randomized Hadamard transform (RHT)...

📖 Read original article


80. Post-Training Language Models for Gold-Medal Performance in Coding Competitions ​

Author: Aleksander Ficek, Sean Narenthiran, Mehrzad Samadi, Somshubra Majumdar, Boris Ginsburg
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.MA, cs.SE

arXiv:2609.02849v1 Announce Type: new Abstract: Competitive programming has become a key test of large language model reasoning, with international competitions such as IOI and ICPC representing its most challenging settings. We present an end-to-end specialization pipeline combining large-scale pro...

📖 Read original article


81. The Implications of Linguistic Illegibility for LLM Security ​

Author: James Mickens
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2609.02852v1 Announce Type: new Abstract: LLMs are trained to generate natural language. However, various strands of evidence indicate that an LLM's externalized linguistic outputs and mechanistically-extracted linguistic features can be an unreliable lens for understanding internal model comp...

📖 Read original article


82. Graph Machine: Towards Better Pretraining via Edges ​

Author: Lintai Hou
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.02881v1 Announce Type: new Abstract: We introduce the Graph Machine (GM), an architecture that maintains an $O(n)$-sized state and accesses it through sparse, dynamic routing. Unlike methods with fixed-size states or sparse but static routing, GM preserves $O(n)$ complexity in its sparse ...

📖 Read original article


83. A Common Measure of Communication for Speech Brain-Computer Interfaces ​

Author: Dulhan Jayalath, Benjamin Ballyk, Oiwi Parker Jones
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC

arXiv:2609.02887v1 Announce Type: new Abstract: Speech brain-computer interfaces (speech BCIs) translate neural activity into language, offering a path towards restoring speech for people with paralysis and, more broadly, enabling new forms of natural human-computer interaction. Despite this promise...

📖 Read original article


Author: Harish Saragadam, Sudhanshu Sharma, Meghana Pujari
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2609.01617v1 Announce Type: cross Abstract: Getting accurate, grounded answers out of large enterprise document repositories is a difficult problem. Dense vector retrieval alone frequently performs poorly on queries that mix technical terminology, vendor-specific acronyms, or require reasoning...

📖 Read original article


85. Multi-Agent Retrieval-Augmented Generation for Efficient Cloud Knowledge Base Search in Telecom SNOC Environment ​

Author: Harish Saragadam, Sudhanshu Sharma, Ipsha Routray
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2609.01618v1 Announce Type: cross Abstract: Telecom Service and Network Operations Centers (SNOCs) rely on large collections of cloud documents, including Standard Operating Procedures (SOPs), vendor technical manuals, incident reports, and configuration guides, to maintain uninterrupted netwo...

📖 Read original article


86. When Literature Data Mislead Artificial Intelligence in Materials Discovery ​

Author: Qian Wang, Ying Li, Ryuhei Sato, Hidemi Kato, Shin-ichi Orimo, Hao Li, Eric Jianfeng Cheng
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IR, cond-mat.mtrl-sci, cs.CE, cs.LG

arXiv:2609.01621v1 Announce Type: cross Abstract: Artificial intelligence (AI) increasingly treats scientific literature as a data source for building databases, training predictive models, and guiding discovery. Yet literature-derived datasets often assume that reported experimental values are inte...

📖 Read original article


87. RecEvolve: A Knowledge-Driven Autonomous Agent System for Recommender Systems ​

Author: Weidi Pan, He Ma, Shuhao Ye, Palaksh Rungta, David McPeek, Junyi Jiao, Arnab Bhadury, Mingyan Gao, Onkar Dalal
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2609.01622v1 Announce Type: cross Abstract: The rise of agentic AI has catalyzed a shift toward self-iterating systems, opening new frontiers for the autonomous optimization of production recommender models. This paper presents the empirical validation of a knowledge-driven autonomous agent sy...

📖 Read original article


88. PRISM: An Agentic Multi-Model Architecture for Proactive Safety in Autonomous Transportation Systems ​

Author: Joyjit Roy, Samaresh Kumar Singh, Sushanta Das
Published: 9/3/2026, 4:00:00 AM
Categories: cs.MA, cs.CV, cs.ET, cs.LG, physics.soc-ph

arXiv:2609.01623v1 Announce Type: cross Abstract: Autonomous and intelligent transportation systems operate in complex urban environments where safety depends on interactions among vehicle behavior, environmental conditions, and vulnerable road users (VRUs) such as pedestrians and cyclists. Most adv...

📖 Read original article


Author: Greg Kocher, Sanjana Arun
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2609.01628v1 Announce Type: cross Abstract: E-commerce search ranking must balance multiple objectives--relevance, user engagement, and platform revenue--when allocating impression slots to competing listings. Estimating the expected revenue component is well understood for fixed-price items, ...

📖 Read original article


90. Omega-N: Interpretable Structural Node Descriptors and Their Applicability Domain ​

Author: Alberto Acedo
Published: 9/3/2026, 4:00:00 AM
Categories: cs.SI, cs.LG, physics.soc-ph, q-bio.MN

arXiv:2609.01633v1 Announce Type: cross Abstract: A composite structural index summarises a network in one number; for a triangle-based index it is spectrally redundant: Tr(A^3) is the third moment of the adjacency spectrum. The non-redundant content sits one level down, in diag(A^3), which depends ...

📖 Read original article


91. SocialBuddy: Tailoring Search Agent for Social Scenarios ​

Author: Mingxuan Li, Yirong Mao, FaZhan Zhang, Haibiao Yao, Runze Hu, Wenhui Que
Published: 9/3/2026, 4:00:00 AM
Categories: cs.SI, cs.LG

arXiv:2609.01641v1 Announce Type: cross Abstract: In the era of digital social interaction, searching friends' posts from massive social streams has become a fundamental user need. However, while modern agentic search frameworks have achieved remarkable success in conventional retrieval tasks, they ...

📖 Read original article


92. Context Inference Attacks Without Jailbreaks ​

Author: Prince Jha, Samuele Poppi, Nils Lukas
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2609.01663v1 Announce Type: cross Abstract: Agentic AI systems are increasingly deployed to process sensitive data at inference time, such as healthcare records or financial documents assembled into a hidden \emph{context} before the system answers. Prior work has studied privacy risks primari...

📖 Read original article


93. Private Computation Space: Experience with Trusted Multi-Cluster Federated Learning for Agriculture ​

Author: Shuangyu Lei, Muhammad Salman Abid, Jacob Belding, Sam Mosher, Manushi B. Trivedi, Shivranjani Baruah, Liam Wickes-Do, Andrew Anderson, Braulio Dumba, Alyssa Whitcraft, Ritvik Sahajpal, Sijin Li, Kelly Robbins, Michael Gore, Margaret Frank, Steven Wolf, Liz Jones, Abraham Stroock, Kaitlin Gold, Hakim Weatherspoon
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2609.01667v1 Announce Type: cross Abstract: Artificial Intelligence has shown to help improve agricultural practices, yet adoption remains limited: 69% of U.S. farmers have privacy concerns with sharing their data, and these concerns must be addressed before adoption is widespread. While Feder...

📖 Read original article


94. Random Forest-Informed Cellular Automaton for Large-Scale Wildfire Spread Modelling ​

Author: Siyu Chen, Esha Saha, Hao Wang
Published: 9/3/2026, 4:00:00 AM
Categories: stat.AP, cs.LG

arXiv:2609.01675v1 Announce Type: cross Abstract: Accurate large-scale wildfire spread modelling requires models that capture both the environmental conditions associated with fire occurrence and the local dynamics of fire propagation. We propose a three-stage framework that combines a Random Forest...

📖 Read original article


95. FORGE: Forward-Only Test-Time Adaptation for Integer-Only Vision Models on Microcontrollers ​

Author: Muhammad Rehan, Haider Ali, Muhammad Ali Munir, Moaz Amjad
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AR, cs.LG

arXiv:2609.01683v1 Announce Type: cross Abstract: Vision models deployed on microcontrollers (MCUs) are quantized to integer-only arithmetic and run in inference-only runtimes that do not carry the machinery backpropagation needs: the standard tool for adapting a model to the distribution shift (sen...

📖 Read original article


96. FairLens: Benchmarking Fairness in Vision-Language Models for High-Stakes Decision-Making ​

Author: Vahid Reza Khazaie, Ahmed Y. Radwan, Shaina Raza
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2609.01691v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used to make decisions from visual inputs. We introduce FAIRLENS, a benchmark and evaluation framework for measuring both the fairness and the validity of VLM responses in three high-stakes domains: hiri...

📖 Read original article


97. Hearing the Whispers: Black-Box Membership Inference Attacks on Finetuned TTS Models ​

Author: Kunlin Cai, Kaiyuan Zhang, Zihang Xiang, Jinghuai Zhang, Abeer Alwan, Fnu Suya, Yuan Tian
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.SD, eess.AS

arXiv:2609.01723v1 Announce Type: cross Abstract: Text-to-Speech (TTS) foundation models are increasingly fine-tuned on private datasets to synthesize highly personalized voices, introducing severe privacy risks by exposing both biometric identities and sensitive speech content. Existing black-box m...

📖 Read original article


98. Harness Engineering in LLM Tool Use via Agent-Native Reusable Tool Primitives ​

Author: Haibo Jin, Suijin Wang, Xucheng Yu, Haojing Luo, Haohan Wang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2609.01736v1 Announce Type: cross Abstract: Large language models (LLMs) augmented with external tools have demonstrated remarkable capability in solving complex real-world tasks. However, existing approaches suffer from two key challenges: brittle multi-step and multi-turn reasoning caused by...

📖 Read original article


99. Pooling and Drift in Delayed Bandits ​

Author: Melika Baghi
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2609.01761v1 Announce Type: cross Abstract: A system often has to act long before it learns whether the act worked: a recommender sees a click in seconds and a purchase in days. With $K$ actions and a delay of $d$ rounds, the best rate known for this setting is $\widetilde{O}(\sqrt{(K+d)T})$ o...

📖 Read original article


100. Ten Architectures, One Error: Shared Failure Modes in Hyperspectral Classification under Spatially Disjoint Evaluation ​

Author: Ehsan Faghih, Fatemeh Ashrafi, Marguerite Moore, Zahra Saki
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2609.01786v1 Announce Type: cross Abstract: Hyperspectral image classification still relies heavily on random pixel splits within a single scene. The Salinas dataset, randomly split, is among the most widely used datasets for comparing different architectures. However, under a random split met...

📖 Read original article


101. Reinforcement learning to choose optimizers ​

Author: Martin van der Schelling, Deepesh Toshniwal, Miguel A. Bessa
Published: 9/3/2026, 4:00:00 AM
Categories: cs.NE, cs.LG, math.OC

arXiv:2609.01811v1 Announce Type: cross Abstract: No single optimization method is uniformly best for all problems, and the most suitable optimizer choice can change during a run. Existing approaches that change optimizer during execution typically predetermine part of the strategy: the portfolio is...

📖 Read original article


102. Interpretable Symptom Vectors for Depression in a Large Language Model ​

Author: Fangyi Zhu, Ajay Subramanian, Allison Constant, Camille Wang, Ravish Gupta, Corey J. Keller
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, q-bio.NC

arXiv:2609.01832v1 Announce Type: cross Abstract: Patients with depression present with diverse symptom profiles, yet clinical practice routinely reduces this variation to a single severity score. Large language models (LLMs) can potentially capture various symptoms and their severity from patient s...

📖 Read original article


103. Latent unified smooth Hamiltonians for excited state chemistry ​

Author: David Juergens, Martin St"ohr, Andreas E. Hillers-Bendtsen, O. Jonathan Fajen, Todd J. Mart'inez
Published: 9/3/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG

arXiv:2609.01871v1 Announce Type: cross Abstract: We describe a neural network architecture and training procedure designed to model electronic ground and excited states of arbitrary molecular systems. By indirectly learning a latent, implicit basis representation of the electronic-state Hamiltonian...

📖 Read original article


104. Basin Geometry and Reliable Recall of Dynamical Memories in Reservoir Computing ​

Author: Ling-Wei Kong, Ying-Cheng Lai
Published: 9/3/2026, 4:00:00 AM
Categories: nlin.CD, cs.LG

arXiv:2609.01914v1 Announce Type: cross Abstract: Reliable attractor recall conventionally requires broad basins of attraction. However, in reservoir-computing based associative memory, temporal cues reliably recover dynamical memories despite basins dominated by unpredictable, riddled-like regions....

📖 Read original article


105. Sparse Readout Prism: Explaining Logit-Lens Scores in Features Instead of Tokens ​

Author: Matteo He, William F. Shen, Xinchi Qiu, Nicholas D. Lane
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2609.01936v1 Announce Type: cross Abstract: A language model's prediction of its next token develops across layers, and lens methods track this process by decoding intermediate hidden states into tokens. But a lens reading reflects both the hidden state and the readout (the unembedding matrix)...

📖 Read original article


106. Pushing Forward Multi-Secret-Key Homomorphic Encryption for Private Average Aggregation ​

Author: Miguel Morona-M'inguez, Fernando P'erez-Gonz'alez, Alberto Pedrouzo-Ulloa
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2609.01945v1 Announce Type: cross Abstract: Federated Learning enables multiple clients to train a shared model while keeping their local datasets isolated. However, the exchanged model updates may still leak sensitive information, making private aggregation a central building block in practic...

📖 Read original article


107. Network-Aware Forecasting on Wireless Access Points ​

Author: Niloo Bahadori, Swadhin Pradhan, Peiman Amini
Published: 9/3/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2609.01957v1 Announce Type: cross Abstract: Enterprise wireless access points (APs) are promising platforms for predictive machine learning (ML), but their primary responsibility remains providing wireless connectivity and network services. Predictive inference must therefore share an AP's CPU...

📖 Read original article


108. Post-Training Ternarization of Qwen3-4B Capability, Effective Bit Budget, Storage Compression, and Deployment ​

Author: Anirudh Malik, M Sparsh Mehra, Poojith Devan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2609.01962v1 Announce Type: cross Abstract: Ultra-low-bit language models can reduce storage and memory bandwidth, but a nominal "1.58-bit" label does not fully describe the stored representation, retained capability, or runtime behavior. We study an end-to-end post-training conversion of Qwen...

📖 Read original article


109. Morphology signal in whole slide image foundation models can automatically triage slides ​

Author: Ayushi Sinha, Shashank Yadav, Benjamin Holmes, Pravat Das, Aaron W. Bogan, James S. Lewis Jr., Santiago Romero-Brufau, Andrew Y. K. Foong, Scott H. Kaufmann, Kathryn M. Van Abel, David M. Routman, Michael R. Lucas
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2609.01987v1 Announce Type: cross Abstract: Patient exams in the cancer diagnosis and staging process typically generate several whole slide images (WSIs). One of the initial steps in training models on WSI data is identifying one or a few slides containing tumor or other diagnostic biomarkers...

📖 Read original article


110. Linear Fusion MultiDiffusion for Fast Training-Free Spherical Panorama Generation ​

Author: Akio Hayakawa, Yusuke Mukuta, Tatsuya Harada
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2609.01997v1 Announce Type: cross Abstract: We propose LF-MultiDiffusion, a training-free panorama generation method that extends MultiDiffusion to support linear projections between target and reference image spaces. Our key idea is to reformulate latent aggregation as a regularized least-squ...

📖 Read original article


111. Posterior Tempering Explains Variance Inflation in Linear and Generalized Linear Thompson Sampling ​

Author: Prateek Jaiswal, Debdeep Pati, Anirban Bhattacharya, Bani K. Mallick
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT, math.ST, stat.TH

arXiv:2609.01999v1 Announce Type: cross Abstract: We study a variant of the Thompson Sampling (TS) algorithm, called $\alpha$-TS, for solving stochastic generalized linear bandit problems. Existing analyses of TS require inflating the posterior variance to derive near-optimal regret guarantees. We f...

📖 Read original article


112. Perceptually Regularized Diffusion Model for Image Super-Resolution ​

Author: Chuxiangbo Wang, Pavithra Venkatachalapathy, Ying Liang, Min Wang, Jing Qin, Yifei Lou, Weihong Guo
Published: 9/3/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2609.02016v1 Announce Type: cross Abstract: Image super-resolution, which aims to reconstruct high-resolution images from their low-resolution observations, is fundamental to medical imaging, remote sensing, surveillance, microscopy, and scientific visualization. Traditional model-based method...

📖 Read original article


113. DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents ​

Author: Zhuoran Yu, Le Thien Phuc Nguyen, Jaden Park, Xinyi Gu, Zexue He, Soochahn Lee, Rogerio Feris, Yong Jae Lee
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG

arXiv:2609.02059v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved strong performance on structured visual understanding tasks such as chart and document question answering. However, existing benchmarks typically evaluate these domains in isolation, leaving unde...

📖 Read original article


114. IDEEA: training-free Input-Dependent stEEring via Activation cluster matching ​

Author: Zheng Wang, Muchen Li, Renjie Liao, Yan Leng
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2609.02089v1 Announce Type: cross Abstract: Steering aligns large language models (LLMs) by injecting a bias into selected activations at inference time, offering a far cheaper alternative to weight-update methods such as supervised fine-tuning or reinforcement learning. However, most existing...

📖 Read original article


115. Disease Burden over Skin Tone: Decomposing the Dermatology-AI Generalization Gap ​

Author: Nirajan Kunwor, Sanjaya Poudel, Quoc-Huy Trinh, Jahidul Arafat, Sunil Kumar Gaire
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2609.02111v1 Announce Type: cross Abstract: Dermatology artificial intelligence (AI) models are predominantly trained on light-skinned, cancer-focused image collections, yet they are increasingly proposed for deployment in resource-constrained settings where patients differ from training popul...

📖 Read original article


116. HyperMC: Multi-Fidelity Hyperparameter Tuning for Stochastic Gradient MCMC ​

Author: Ming Tan, Xiyun Jiao
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO

arXiv:2609.02138v1 Announce Type: cross Abstract: Stochastic gradient Markov chain Monte Carlo (SGMCMC) methods enable scalable Bayesian inference, but their performance depends strongly on hyperparameters such as the step size, mini-batch size, and number of leapfrog steps. Since most SGMCMC algori...

📖 Read original article


117. SoK: Where Do Flow Labels Come From? Auditing Label Provenance in Encrypted Traffic Benchmarks ​

Author: Sizhe Huang, Shujie Yang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2609.02140v1 Announce Type: cross Abstract: Encrypted traffic classification infers semantics beyond the flow record from transport-layer observables, and supervised training rests on labels that hold for the individual flow they are attached to. Recent systematizations scrutinize model in- pu...

📖 Read original article


Author: Sajad Faghfoor Maghrebi, Navid Eslami, Niv Dayan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.DB, cs.AI, cs.IR, cs.LG

arXiv:2609.02143v1 Announce Type: cross Abstract: Most vector databases rely on graph-based indexes, notably HNSW and Vamana, for approximate nearest neighbor search. With embedding models widely adopted, the datasets these databases store grow rapidly. At a fixed accuracy, how does search cost scal...

📖 Read original article


119. GenCAR: Generative Counterfactual Alignment with Risk-Controlled Selection for Out-of-Distribution Recommendation ​

Author: Qianqian Wang, Yunshan Li, Jiawen Zeng, Wenwu Gong, Lili Yang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2609.02162v1 Announce Type: cross Abstract: Serving useful recommendations under distribution shift is crucial for balancing utility and risk in out-of-distribution (OOD) recommendation. However, most existing OOD methods improve ranking or construct counterfactual candidates without controlli...

📖 Read original article


Author: Shiliang Xiao, Jingsong Wei, Yuzhi Liang, Yufan Zheng, Xia Li, Qiliang Lin
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2609.02172v1 Announce Type: cross Abstract: Optimization-based jailbreak attacks such as Greedy Coordinate Gradient (GCG) achieve strong effectiveness and transferability by optimizing adversarial suffixes on white-box source models. However, existing GCG-based methods rely on averaged adversa...

📖 Read original article


121. WeaveMark: Robust and Scalable Multi-bit LLM Watermarking via Coded Payload Spreading ​

Author: Gang-Hyun Park, Ju-Hyeong Lee, Hee-Youl Kwak, Dae-Young Yun
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2609.02177v1 Announce Type: cross Abstract: Multi-bit watermarking for large language models (LLMs) enables content source tracing by embedding user-identifiable messages into generated text. Existing methods face a fundamental trade-off among extraction accuracy, text quality, and payload cap...

📖 Read original article


122. Quantum MeanFlow: single-shot generative sampling on NISQ hardware ​

Author: Ashish Joshi, Eshaan Mistry, Takahiko Koyama
Published: 9/3/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2609.02186v1 Announce Type: cross Abstract: Quantum generative models offer a promising framework for exploring whether quantum computation can enhance generative machine learning. Flow matching is a generative method in which samples are generated by transporting a simple, known distribution ...

📖 Read original article


123. Schr\"odinger Bridges on Lie Group Manifolds for Probabilistic Intrinsic Generation ​

Author: Shizhe Zhang, Mingyang Zhao, Lei Ma
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2609.02196v1 Announce Type: cross Abstract: Generative modeling directly on geometric manifolds can avoid errors introduced by flattening non-Euclidean data, repeated ambient projection, and coordinate inconsistency in Euclidean representations. Schrodinger bridges provide a probabilistic gene...

📖 Read original article


124. Prototype-guided transfer of sparse literature knowledge for electrolyte additive discovery ​

Author: Weixiang Hong, Hongting Du, Jiayue Tang, Ruifeng Tan, Yangjian Quan, Jia Li, Jiaqiang Huang
Published: 9/3/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG

arXiv:2609.02209v1 Announce Type: cross Abstract: Electrolyte additive discovery remains challenging because experimentally validated molecules are sparse, whereas accessible chemical spaces are vast and largely unlabeled. This challenge is amplified in lithium-ion batteries, where additive performa...

📖 Read original article


125. Hardware-Accelerated Instance Segmentation for Resource-Constrained Space Robotics with Criticality Analysis ​

Author: Siddhant Shete, Hilmi Dogu K"uc"uker, Udo Frese, Frank Kirchner
Published: 9/3/2026, 4:00:00 AM
Categories: cs.RO, cs.AR, cs.CV, cs.LG

arXiv:2609.02219v1 Announce Type: cross Abstract: Autonomous lunar missions require real-time per- ception under three coupled constraints: extreme low-light conditions, limited onboard compute, and radiation-induced hardware faults that can silently corrupt inference. We present a deployment-orient...

📖 Read original article


126. LLM-as-a-Judge Is Not an Oracle: Why Self-Improving Agents Need Deterministic Guardrails ​

Author: Vansh Wahi
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2609.02246v1 Announce Type: cross Abstract: Self-improving agent pipelines have a problem at their center. An optimizer rewrites prompts to score higher, and the score comes from a judge that is itself an LLM. That judge has the last word on whether the system is getting better, and our positi...

📖 Read original article


127. RideSkill: A Hierarchical Algorithm for Generalized Ride Sharing with LLM-Driven Automatic Evolution ​

Author: Zijian Zhao, Sen Li, Xialiang Tong, Mingxuan Yuan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.MA, cs.CL, cs.ET, cs.LG

arXiv:2609.02250v1 Announce Type: cross Abstract: Ride-sharing, which allows multiple passengers with different origin-destination (OD) pairs to share a single vehicle, is a challenging operational problem, as it requires orders with different OD pairs to be efficiently bundled and assigned to vehic...

📖 Read original article


128. Codebook Agent: Amortized Topology Design for LLM Multi-Agent Systems ​

Author: Jinxi Yu, Yubei Li, Eric Hanchen Jiang, Zhi Zhang, Dong Liu, Wenxiao Zhao, Levina Li, Kai-Wei Chang, Ying Nian Wu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2609.02264v1 Announce Type: cross Abstract: Adapting the communication topology of an LLM multi-agent system to each query improves both accuracy and efficiency, yet current designers treat this as conditional graph generation: a variational, autoregressive, or diffusion decoder searches the $...

📖 Read original article


129. Do Large Language Models Capture the Diversity in their Training Data? ​

Author: Youqi Wu, Farzan Farnia
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2609.02275v1 Announce Type: cross Abstract: Large language models are trained to model conditional distributions over text, yet it remains inadequately understood whether they capture the full diversity of plausible outputs present in their training data. We study this question through an info...

📖 Read original article


130. From topology learning to graph generation: A unifying perspective ​

Author: Xiaowen Dong, Hoi-To Wai, Siheng Chen, Laura Toni, Dorina Thanou
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, eess.SP

arXiv:2609.02286v1 Announce Type: cross Abstract: Learning graph structures from data is a fundamental problem that spans a wide range of signal processing and machine learning tasks. While significant effort has been made to tackle the problem, existing research has largely evolved along two parall...

📖 Read original article


131. Improving Evaluation Realism with Inference-Time Compute and Deployment Scaffolds ​

Author: Axel Ahlqvist, Richard Guan, Juan-Pablo Rivera, Adeline Kassler, Dmitrii Troitskii, Alexandra Souly, Kai Fronsdal, Robert Kirk, John Hughes
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2609.02302v1 Announce Type: cross Abstract: A core obstacle to alignment evaluation is evaluation awareness: capable models can tell when they are being tested rather than deployed, weakening the conclusions a safety evaluation can support. We present two techniques that make simulated alignme...

📖 Read original article


132. Poisoning Attacks on the PGM-index ​

Author: Atsuki Sato, Martin Aum"uller, Yusuke Matsui
Published: 9/3/2026, 4:00:00 AM
Categories: cs.DB, cs.CR, cs.LG

arXiv:2609.02328v1 Announce Type: cross Abstract: The PGM-index (Ferragina and Vinciguerra, VLDB'20) is one of the most practical learned indexes, owing to its theoretical elegance and consistently strong empirical performance. It is built on optimal piecewise linear approximations (PLAs) that minim...

📖 Read original article


133. Humanoid Safe Stop via Learned Stoppability Value ​

Author: Junfeng Long, Pieter Abbeel, Koushil Sreenath, Roberto Horowitz, Guanya Shi, C. Karen Liu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SY, eess.SY

arXiv:2609.02358v1 Announce Type: cross Abstract: Humanoid robots responding to emergency stop commands typically execute a fixed maneuver, without reasoning about whether a safe stop is actually feasible from the current state. We cast emergency stopping as a reach-avoid problem and propose Safe-St...

📖 Read original article


134. A computational approach to maximum likelihood thresholds for colored Gaussian graphical models ​

Author: Roser Homs, Olga Kuznetsova, Bernadette J. Stolz
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.AG, math.ST, stat.TH

arXiv:2609.02382v1 Announce Type: cross Abstract: Gaussian graphical models (GGMs) are essential tools for interpretable structure learning. However, in high-dimensional, small-sample regimes, the available data is often insufficient for the maximum likelihood estimator to exist. Colored Gaussian gr...

📖 Read original article


135. When Decodability Is Not Enough: Logical Validity Representations, Behavioral Dissociation, and Causal Tests in Language Models ​

Author: Smitha Muthya Sudheendra, Jaideep Srivastava
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2609.02438v1 Announce Type: cross Abstract: Large language models can look capable of logical reasoning, but correct or incorrect answers alone tell us little about what the model represents internally. We study logical verification in five open-weight transformer models using matched valid--i...

📖 Read original article


136. Training seeds and model-selection stability in recommender-system evaluation ​

Author: Juan Manuel Rodriguez, Oleg Lesota, Antonela Tommasel
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2609.02499v1 Announce Type: cross Abstract: Recommender-system experiments often rely on a single random training seed, assuming that run-to-run stochasticity has limited impact on evaluation conclusions. This assumption is risky, as a training seed may influence several algorithm-dependent me...

📖 Read original article


137. Orthogonal Ensembles and Tested Explanations for Performer-Independent Body-Motion Emotion Recognition ​

Author: Naoto Nishida, Yoshio Ishiguro
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.HC, cs.LG

arXiv:2609.02510v1 Announce Type: cross Abstract: We study body-only, 12-class acted-emotion classification from skeleton motion under leave-performer-out (LPO) evaluation, a hard, underdetermined setting: chance is 8.3%, and a protocol-matched reproduced STGCN++ baseline reaches only 25.73 +/- 4.03...

📖 Read original article


138. Learning-Based Reconstruction Attacks on Coordinate-Obfuscated Point Clouds ​

Author: Mohammad Waquas Usmani, Susmit Shannigrahi, Michael Zink
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2609.02568v1 Announce Type: cross Abstract: Volumetric video based on point cloud representations enables immersive virtual and augmented reality applications but introduces significant challenges for efficient and secure content delivery. Prior work proposed a selective coordinate encryption ...

📖 Read original article


139. Scalable Direction-Following TTS via Voice Impression-Guided Pseudo Triplet Construction ​

Author: Kenichi Fujita, Yusuke Ijima
Published: 9/3/2026, 4:00:00 AM
Categories: cs.SD, cs.CL, cs.LG

arXiv:2609.02623v1 Announce Type: cross Abstract: Voice actors often re-read the same script while modifying their delivery in response to performance directions. We study this setting as direction-following TTS, where a system generates a new utterance that reflects a given direction relative to a ...

📖 Read original article


140. TaRA: Training-Aware Low-Rank Adaptation Initialization ​

Author: Taehyeon Kim, Eunhyeok Park
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2609.02639v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become a de facto standard for parameter-efficient fine-tuning (PEFT), yet its performance is highly sensitive to initialization due to the information bottleneck imposed by low-rank decomposition. Existing approaches a...

📖 Read original article


141. Loom: Weaving Diagnostic Strands into Free-Text Consensus via Embedding-Space Reweighting ​

Author: Ron Begleiter, Katya Egert Berg, Gilad Saban, Gil Shabat
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2609.02649v1 Announce Type: cross Abstract: Aggregating noisy, conflicting textual hypotheses into a reliable consensus is a fundamental challenge when deploying NLP systems in real-world industrial settings. While monolithic Large Language Model (LLM) agents offer unbounded expressivity for t...

📖 Read original article


142. Dimension Dependent Correlation Gap Bounds under Restricted Independence ​

Author: Arjun Ramachandra
Published: 9/3/2026, 4:00:00 AM
Categories: math.PR, cs.LG, math.CO

arXiv:2609.02659v1 Announce Type: cross Abstract: The pairwise independent correlation gap is the ratio of the maximum expected value of a set function under arbitrary dependence to that under pairwise independence, measuring the loss from this independence restriction. Under mutual independence, th...

📖 Read original article


143. oHC: Orthogonal Hyper-Connections on SO(4) via Quaternions ​

Author: Haoqiang Guo, Xuyi Chen, Bo Ke, Yishu Lei, Ziyang Xu, Shikun Feng, Ximen, Wenhan Luo
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2609.02672v1 Announce Type: cross Abstract: Hyper-Connections (HC) replace the single residual stream of a Transformer with $n$ parallel ones, mixing them at every layer with a learned $n \times n$ residual matrix. Leaving that matrix unconstrained places no limit on the factor by which the mi...

📖 Read original article


144. Eliciting ESG Preferences for Reinforcement Learning-Based Portfolio Optimization ​

Author: Giovanni Dispoto, Marcello Restelli, Carmine Ventre
Published: 9/3/2026, 4:00:00 AM
Categories: q-fin.PM, cs.CE, cs.LG

arXiv:2609.02677v1 Announce Type: cross Abstract: Modern portfolio management increasingly demands a balance between traditional risk-adjusted returns and strict Environmental, Social, and Governance (ESG) mandates. Current Reinforcement Learning (RL) approaches typically optimize for a single ESG p...

📖 Read original article


145. Neural operators approximate strongly continuous convex monotone semigroups ​

Author: Jonas Blessing, Philipp Schmocker, Alessandro Sgarabottolo
Published: 9/3/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.AP, math.PR, stat.ML

arXiv:2609.02727v1 Announce Type: cross Abstract: We approximate strongly continuous convex monotone semigroups by learning their Chernoff-type one-step operators with neural operators. First, we introduce the general class of so-called Chernoff-neural operators and show in a universal approximation...

📖 Read original article


146. Momentum in large-batch training: Polyak enlarges the critical batch size, Nesterov improves data efficiency ​

Author: Jia-Nan Wang, Zixun Huang, Kairui Li, Lei Wu
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC

arXiv:2609.02728v1 Announce Type: cross Abstract: We study when and how momentum improves large-batch training in the one-pass regime, using power-law kernel regression as a tractable setting. We first characterize risk stability through the critical learning rate, defined as the largest learning ra...

📖 Read original article


147. Language Models Can Control Their Own Attention ​

Author: Namgyu Ho, Huzama Ahmad, Woosung Koh, Se-Young Yun, Tal Schuster, Cicero Nogueira dos Santos
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2609.02737v1 Announce Type: cross Abstract: Language models spend most of their attention on a small fraction of context, yet they read the entire KV cache to find the few tokens that matter. If the user asks about a previous detail in a 1M-token conversation, global attention layers must scan...

📖 Read original article


148. SPADE: SPaT Attack Detection from the Connected Vehicle's Perspective ​

Author: James Di Novo, Hany Ragab, Sylvain P. Leblanc
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2609.02741v1 Announce Type: cross Abstract: Signal Phase and Timing (SPaT) messages are a cornerstone of connected vehicle (CV) safety, enabling CVs to perceive and respond to intersection state through Vehicle-to-Infrastructure (V2I) and Vehicle-to-Vehicle (V2V) communication. The integrity o...

📖 Read original article


149. HiPoly: a hierarchical polymer-native AI framework for property prediction and generative design ​

Author: Ge Sun, Gervasio Zaldivar, Yuan Tian, Gustavo Perez Lemus, Juhae Park, Dasha Safarian, Ming Han, Juan J. de Pablo
Published: 9/3/2026, 4:00:00 AM
Categories: physics.chem-ph, cond-mat.mtrl-sci, cs.AI, cs.LG, physics.comp-ph

arXiv:2609.02746v1 Announce Type: cross Abstract: Polymeric materials are central to modern technologies, with applications ranging from energy to health and transportation. Although AI has made significant advances in materials discovery, the hierarchical structure of polymers across multiple lengt...

📖 Read original article


150. Untangling the Mechanisms of Misleading Context in Medical Question Answering ​

Author: Robin Linzmayer, No'emie Elhadad
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2609.02754v1 Announce Type: cross Abstract: Large language models now answer medical questions with expert-level performance. However, the context these systems act on can be misleading, and misleading context can corrupt a model's medical judgment. To understand how misleading context corrupt...

📖 Read original article


151. From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution ​

Author: Yuzhang Luo, Chenpeng Wang, Jianhui Chen, Liangming Pan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2609.02771v1 Announce Type: cross Abstract: Training data attribution (TDA) aims to identify training examples that shape model behavior, but its intervention value depends on both which examples are selected and how they are modified. Influence functions (IF) estimate behavioral changes under...

📖 Read original article


152. CodePoisonRAG: Knowledge Poisoning Attacks on Retrieval-Augmented Code Generation ​

Author: Varun Gadey, Ziad Marey, Alexandra Dmitrienko
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2609.02774v1 Announce Type: cross Abstract: Retrieval-Augmented Code Generation (RACG) improves LLM-based software development by retrieving external code artifacts, documentation, and patches, and incorporating them into the generation context. This reliance on external knowledge introduces a...

📖 Read original article


153. Full-Model Optimality for Tunable Linear Generative Priors in Compressed Sensing ​

Author: Zhaoming Li, Paul Hand
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2609.02790v1 Announce Type: cross Abstract: Generative models have been studied experimentally and theoretically as priors for inverse problems such as compressed sensing. Recent work by Gunn et al. studied the use of generative priors with tunable complexity, where a family of generative prio...

📖 Read original article


154. Dutch Books for Language Models ​

Author: Isaiah Andrews, Suproteem Sarkar
Published: 9/3/2026, 4:00:00 AM
Categories: econ.GN, cs.AI, cs.CL, cs.LG, q-fin.EC

arXiv:2609.02797v1 Announce Type: cross Abstract: People increasingly use language models to support life decisions. Many such decisions involve a probabilistic forecast: How likely is a major life event, a natural disaster, or an economic outcome? Users of language models may implicitly trust that ...

📖 Read original article


155. AI Contextual Measurement for Recovering Individual and Group-Level Effects: Validation Against Survey Measures and an Occupational Application ​

Author: Wenxin Jiang, Xuyang Wang, Yuxiao Wu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2609.02821v1 Announce Type: cross Abstract: Researchers increasingly use artificial intelligence to construct measures of social, organizational, and occupational characteristics that are absent from conventional surveys. We propose AICOME, AI COntextual MEasurement, a framework for evaluating...

📖 Read original article


156. Learning Spectral-Like Mesh-Free Discretisations ​

Author: Lucas Gerken Starepravo, Henry Broadley, Steven Lind, Jack R. C. King
Published: 9/3/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG

arXiv:2609.02833v1 Announce Type: cross Abstract: Meshfree methods such as smoothed particle hydrodynamics (SPH) with kernel corrections, radial basis function-generated finite differences (RBF-FD), and the local anisotropic basis function method (LABFM) construct discrete differential operators by ...

📖 Read original article


157. Improved Gradient Descent Lower Bounds Beyond Nesterov ​

Author: Yuhan Ye, Kaizhao Liu
Published: 9/3/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ML

arXiv:2609.02855v2 Announce Type: cross Abstract: We study how far gradient descent (GD) can be accelerated by predetermined stepsizes in smooth convex optimization. Going beyond the classical $\Omega(n^{-2})$ first-order oracle lower bound of Nemirovsky and Yudin (1983), we prove an $\Omega(n^{-1.6...

📖 Read original article


158. GRADSOLVE: fast exact gradients for ODE ensembles on GPUs ​

Author: Alessio Spurio Mancini
Published: 9/3/2026, 4:00:00 AM
Categories: cs.MS, cs.DC, cs.LG, cs.NA, math.NA

arXiv:2609.02876v1 Announce Type: cross Abstract: Ordinary differential equations (ODEs) underlie models in science and engineering, and many applications need derivatives of their solutions with respect to parameters. Ensembles of independent trajectories suit graphics processing units (GPUs), but ...

📖 Read original article


159. Discriminative World Models for Web Agents ​

Author: Kelvin Li, Dhruv Pendharkar, Anish Pahilajani, Chuyi Shang, Leon Oks, Leonid Karlinsky, Rogerio Feris, Trevor Darrell, Roei Herzig
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2609.02885v1 Announce Type: cross Abstract: Recent web agents use world models for test-time action selection by sampling candidate actions, predicting the resulting web states, and ranking them with a ranker model or a Process Reward Model (PRM). These world models are typically trained via s...

📖 Read original article


160. Deep denoising autoencoder-based non-invasive blood flow detection for arteriovenous fistula ​

Author: Li-Chin Chen, Yi-Heng Lin, Li-Ning Peng, Feng-Ming Wang, Yu-Hsin Chen, Po-Hsun Huang, Shang-Feng Yang, Yu Tsao
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP

arXiv:2306.06865v2 Announce Type: replace Abstract: Clinical guidelines underscore the importance of regularly monitoring and surveilling arteriovenous fistula (AVF) access in hemodialysis patients to promptly detect any dysfunction. Although phono-angiography/sound analysis overcomes the limitation...

📖 Read original article


161. Gradient Descent on Logistic Regression with Non-Separable Data and Large Step Sizes ​

Author: Si Yi Meng, Antonio Orvieto, Daniel Yiming Cao, Christopher De Sa
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2406.05033v3 Announce Type: replace Abstract: We study gradient descent (GD) dynamics on logistic regression problems with large, constant step sizes. For linearly-separable data, it is known that GD converges to the minimizer with arbitrarily large step sizes, a property which no longer holds...

📖 Read original article


162. Smoothed Analysis for Learning Concepts with Low Intrinsic Dimension ​

Author: Gautam Chandrasekaran, Adam Klivans, Vasilis Kontonis, Raghu Meka, Konstantinos Stavropoulos
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CC

arXiv:2407.00966v3 Announce Type: replace Abstract: In traditional models of supervised learning, the goal of a learner-- given examples from an arbitrary joint distribution on $\mathbb{R}^d \times {\pm 1}$-- is to output a hypothesis that is competitive (to within $\epsilon$) of the best fitting ...

📖 Read original article


163. Prompting the Unknown: Understanding Response Uncertainty in Large Language Models ​

Author: Ze Yu Zhang, Arun Verma, Finale Doshi-Velez, Bryan Kian Hsiang Low
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2407.14845v4 Announce Type: replace Abstract: Large language models (LLMs) are widely used in decision-making across diverse domains. Ensuring the generation of safe and reliable responses is critical for the effective deployment of LLM-based applications, particularly in high-stakes domains s...

📖 Read original article


164. Doubly Stochastic Adaptive Neighbors Clustering via the Marcus Mapping ​

Author: Jinghui Yuan, Chusheng Zeng, Fangyuan Xie, Zhe Cao, Mulin Chen, Rong Wang, Feiping Nie, Yuan Yuan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2408.02932v3 Announce Type: replace Abstract: Clustering is a fundamental task in machine learning and data science, and similarity graph-based clustering is an important approach within this domain. Doubly stochastic symmetric similarity graphs provide numerous benefits for clustering problem...

📖 Read original article


165. Achieving More with Less: A Tensor-Optimization-Powered Ensemble Method ​

Author: Jinghui Yuan, Weijin Jiang, Zhe Cao, Fangyuan Xie, Rong Wang, Feiping Nie, Yuan Yuan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2408.02936v3 Announce Type: replace Abstract: Ensemble learning is a method that leverages weak learners to produce a strong learner. However, obtaining a large number of base learners requires substantial time and computational resources. Therefore, it is meaningful to study how to achieve th...

📖 Read original article


166. Action abstractions for amortized sampling ​

Author: Oussama Boussif, L'ena N'ehale Ezzine, Joseph D Viviano, Micha{\l} Koziarski, Moksh Jain, Esmeralda S. Whitammer, Emmanuel Bengio, Rim Assouel, Yoshua Bengio
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2410.15184v2 Announce Type: replace Abstract: As trajectories sampled by policies used by reinforcement learning (RL) and generative flow networks (GFlowNets) grow longer, credit assignment and exploration become more challenging, and the long planning horizon hinders mode discovery and genera...

📖 Read original article


167. Monotonic anomaly detection ​

Author: Oliver Urs Lenz, Matthijs van Leeuwen
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2410.23158v3 Announce Type: replace Abstract: Semi-supervised anomaly detection is based on the principle that any record that looks different from normal training data is a potential anomaly. However, in some cases we are specifically interested in anomalies that correspond to high attribute ...

📖 Read original article


168. Double-Bounded Nonlinear Optimal Transport for Size Constrained Min Cut Clusterin ​

Author: Fangyuan Xie, Jinghui Yuan, Feiping Nie, Xuelong Li
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2501.18143v2 Announce Type: replace Abstract: Min cut is an important graph partitioning method. However, current solutions to the min cut problem suffer from slow speeds, difficulty in solving, and often converge to simple solutions. To address these issues, we relax the min cut problem into ...

📖 Read original article


169. Nonasymptotic CLT and Error Bounds for Linear Two-Time-Scale Stochastic Approximation ​

Author: Seo Taek Kong, Sihan Zeng, Thinh T. Doan, R. Srikant
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2502.09884v4 Announce Type: replace Abstract: We consider linear two-time-scale stochastic approximation algorithms driven by martingale noise. Recent applications in machine learning motivate the need to understand finite-time error rates, but conventional stochastic approximation analyses fo...

📖 Read original article


170. No Data Wasted: A Semi-supervised Generative Model for Incomplete Multi-view Data Integration with Missing Labels ​

Author: Yiyang Shen, Weiran Wang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2508.11180v2 Announce Type: replace Abstract: Multi-view learning is widely applied to real-life datasets, but it often suffers from both missing views and missing labels. Prior probabilistic approaches addressed the missing view problem by using a product-of-experts scheme to aggregate repres...

📖 Read original article


171. Simulating Classification Models for Ex-Ante Evaluation of Predict-Then-Optimize Methods ​

Author: Pieter Smet
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.02191v3 Announce Type: replace Abstract: Predict-Then-Optimize combines machine learning predictions with downstream optimization to support decision-making when problem parameters are unknown at the time of solving. However, better predictive performance does not necessarily lead to bett...

📖 Read original article


172. General Demographic Pre-trained Models for Enhancing Predictive Performance Across Diseases and Population ​

Author: Li-Chin Chen, Ji-Tian Sheu, Yuh-Jue Chuang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2509.07330v3 Announce Type: replace Abstract: Foundation models for healthcare require balancing robust generalization across heterogeneous clinical populations and disease settings with the architectural simplicity needed for deployment. We present a pre-trained model focused on demographic a...

📖 Read original article


173. Exchange Policy Optimization Algorithm for Semi-Infinite Safe Reinforcement Learning ​

Author: Jiaming Zhang, Yujie Yang, Haoning Wang, Liping Zhang, Shengbo Eben Li
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.04147v2 Announce Type: replace Abstract: Safe reinforcement learning (RL) aims to optimize long-term performance while adhering to safety requirements. However, many practical applications involve an infinite number of constraints, forming semi-infinite safe RL (SI-safe RL). Such scenario...

📖 Read original article


174. Gradient Prediction with Control Variates in the Cheap-Forward Regime ​

Author: Kamil Ciosek, Nicol`o Felicioni, Juan Elenter, Ehsan Imani
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2511.05187v2 Announce Type: replace Abstract: We study whether otherwise-idle inference resources could reduce the scarce-GPU cost of training. Our analysis uses a simulated compute ledger in which fleet work is billed at a fraction of a scarce-GPU forward; all experiments run on a regular GPU...

📖 Read original article


175. SEBA: Sample-Efficient Black-Box Attacks on Visual Reinforcement Learning ​

Author: Tairan Huang, Yulin Jin, Junxu Liu, Qingqing Ye, Haibo Hu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2511.09681v3 Announce Type: replace Abstract: Visual reinforcement learning has achieved remarkable progress in visual control and robotics, but its vulnerability to adversarial perturbations remains underexplored. Most existing black-box attacks focus on vector-based or discrete-action RL, an...

📖 Read original article


176. Freeze, Diffuse, Decode: Task-Aware Adaptation of Transformer Embeddings for Antimicrobial Peptide Design ​

Author: Pankhil Gawade, Adam Izdebski, Myriam Lizotte, Kevin R. Moon, Jake S. Rhodes, Guy Wolf, Ewa Szczurek
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.23120v3 Announce Type: replace Abstract: Pretrained transformers provide rich, general-purpose embeddings, which are transferred to downstream tasks. However, current transfer strategies: fine-tuning and probing, either distort the pretrained geometric structure of the embeddings or lack ...

📖 Read original article


177. A Multivariate Bernoulli-Based Sampling Method for Multi-Label Data with Application to Meta-Research ​

Author: Simon Chung, Colby J. Vorland, Donna L. Maney, Andrew W. Brown
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2512.08371v5 Announce Type: replace Abstract: Datasets may contain observations with multiple labels. If the labels are not mutually exclusive, and if the labels vary greatly in frequency, obtaining a sample that includes sufficient observations with scarcer labels to make inferences about tho...

📖 Read original article


178. Cantelli Constrained Policy Optimization ​

Author: Rohan Tangri, Jan-Peter Calliess
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2601.22993v5 Announce Type: replace Abstract: We introduce Canary, a risk-averse method designed to optimize Value-at-Risk (VaR) constrained reinforcement learning (RL) problems. We employ Cantelli's inequality to obtain a tractable, conservative and smooth bound on the VaR constraint based on...

📖 Read original article


179. Shiva-DiT: Residual-Based Differentiable Top-$k$ Selection for Efficient Diffusion Transformers ​

Author: Jiaji Zhang, Hailiang Zhao, Jiaju Wu, Ruichao Sun, Xinkui Zhao, Shuiguang Deng
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2602.05605v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) are costly at high resolution because self-attention scales quadratically with token sequence length. Existing pruning methods do not jointly provide end-to-end learnability, low training overhead, and deterministic to...

📖 Read original article


180. Constrained Group Relative Policy Optimization ​

Author: Roger Girgis, Rodrigue de Schaetzen, Luke Rowe, Azal'ee Robitaille, Christopher Pal, Liam Paull
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.RO

arXiv:2602.05863v4 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) remains the dominant critic-free approach for fine-tuning LLMs and VLMs, but its compatibility with constrained policy optimization (e.g. for safety-critical domains) has not been carefully examined. In thi...

📖 Read original article


181. MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models ​

Author: Chen-Hao Chao, Wei-Fang Sun, Junwei Quan, Chun-Yi Lee, Rahul G. Krishnan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.16077v4 Announce Type: replace Abstract: Masked diffusion models (MDM) exhibit superior generalization when learned using a Partial masking scheme (Prime). This approach converts tokens into sub-tokens and models the diffusion process at the sub-token level. We identify two limitations of...

📖 Read original article


182. MISApp: Multi-Hop Intent-Aware Session Graph Learning for Next App Prediction ​

Author: Yunchi Yang, Longlong Li, Jianliang Wu, Cunquan Qu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.21653v2 Announce Type: replace Abstract: Predicting the next mobile app a user will launch is essential for proactive mobile services. Yet accurate prediction remains challenging in real-world settings, where user intent can shift rapidly within short sessions and user-specific historical...

📖 Read original article


183. SpecXMaster Technical Report ​

Author: Yutang Ge, Yaning Cui, Hanzheng Li, Jun-Jie Wang, Fanjie Xu, Jinhan Dong, Yongqi Jin, Dongxu Cui, Peng Jin, Guojiang Zhao, Hengxing Cai, Tianci Yangfeng, Xueqing Chen, Hongshuai Wang, Rong Zhu, Linfeng Zhang, Xiaohong Ji, Zhifeng Gao
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.23101v4 Announce Type: replace Abstract: Intelligent spectroscopy serves as a pivotal element in AI-driven closed-loop scientific discovery, functioning as the critical bridge between matter structure and artificial intelligence. However, conventional expert-dependent spectral interpretat...

📖 Read original article


184. Train at Moving Edge: Online-Verified Prompt Selection for Efficient RL Training of Large Reasoning Model ​

Author: Jiahao Wu, Ning Lu, Shengcai Liu, Kun Wang, Yanting Yang, Baijiong Lin, Chen Jason Zhang, Li Qing, Ke Tang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.25184v3 Announce Type: replace Abstract: Reinforcement learning (RL) has become essential for post-training large language models (LLMs) in reasoning tasks. While scaling rollouts can stabilize training and enhance performance, the computational overhead is a critical issue. In algorithms...

📖 Read original article


185. Beyond State Consistency: Behavior Consistency in Text-Based World Models ​

Author: Youling Huang, Guanqiao Chen, Junchi Yao, Lu Wang, Fangkai Yang, Chao Du, ChenZhuo Zhao, Pu Zhao, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.13824v2 Announce Type: replace Abstract: World models have been emerging as critical components for assessing the consequences of actions generated by interactive agents in online planning and offline evaluation. In text-based environments, world models are typically evaluated and trained...

📖 Read original article


186. On the Expressive Power and Limitations of Multi-Layer SSMs ​

Author: Nikola Zubi'c, Qian Li, Yuyi Wang, Davide Scaramuzza
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CC

arXiv:2604.14501v2 Announce Type: replace Abstract: We study how depth, finite precision, state dimension, and chain-of-thought (CoT) affect the expressive power of multi-layer state-space models (SSMs). For the explicit-table $K$-function-composition problem, a canonical benchmark for sequential in...

📖 Read original article


187. Stream-CQSA: Exact Out-of-Memory Recovery for Attention ​

Author: Yiming Bian, Joshua M. Akey
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2604.20819v2 Announce Type: replace Abstract: Long-context large language models are limited not only by attention cost but also by out-of-memory (OOM) failures. A selected attention call may not fit in available device memory even when the kernel is optimized. Exact and approximate attention ...

📖 Read original article


188. RCProb: Probabilistic rule extraction from classification tree ensembles ​

Author: Josue Obregon
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.25304v3 Announce Type: replace Abstract: Tree ensembles provide strong classification performance but usually behave as black-box models. Post-hoc interpretability techniques such as RuleCOSI+ extract a small ruleset that approximates the ensemble, but this simplification can leave the pr...

📖 Read original article


189. Language Diffusion Models are Associative Memories Capable of Retrieving Unseen Data ​

Author: Bao Pham, Mohammed J. Zaki, Luca Ambrogioni, Dmitry Krotov, Matteo Negri
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2604.26841v2 Announce Type: replace Abstract: When do language diffusion models memorize their training data, and how to quantitatively assess their true generative regime? We address these questions by showing that Uniform-based Discrete Diffusion Models (UDDMs) fundamentally behave as Associ...

📖 Read original article


190. When Prompts Interact: Assessing Prompt Arithmetic for Deconfounding under Distribution Shift ​

Author: Zhecheng Sheng, Yongsen Tan, Xiruo Ding, Trevor Cohen, Serguei Pakhomov
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2605.03096v2 Announce Type: replace Abstract: In classification tasks, models may rely on confounding variables to achieve strong in-distribution performance, capturing spurious features that fail under distribution shift. This shortcut behavior leads to substantial degradation in out-of-distr...

📖 Read original article


191. Inference-Native Zeroth-Order Optimization ​

Author: Zelin Li, Caiwen Ding
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.28760v2 Announce Type: replace Abstract: Zeroth-order (ZO) optimization removes backpropagation, but conventional implementations still create candidate states by mutating model weights and materialize updates through the full parameter state. We introduce Inference-Native ZO, which expos...

📖 Read original article


192. Shortcomings and capacities of real-constrained neural networks in complex spaces ​

Author: Andrew Gracyk
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, math.PR

arXiv:2606.04390v3 Announce Type: replace Abstract: We find the asymptotic ratio between the storage capacities when enforcing real pre-activations in a complex hypothesis class as opposed to complex ones in the same class. We use weights drawn from the complex Gaussian, which converge asymptoticall...

📖 Read original article


193. Enabling KV Caching of Shared Prefix for Diffusion Language Models ​

Author: Younghun Go, Jaehoon Han, Changyong Shin, Chuck Yoo, Gyeongsik Yang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.07571v4 Announce Type: replace Abstract: Key-value (KV) caching for shared prefixes is essential for high-throughput large language model (LLM) serving, but it faces critical challenges in emerging diffusion language models (DLMs). In DLMs, bidirectional attention means that updating any ...

📖 Read original article


194. DOG-DPO:Dynamic Optimization in Geometry for Safety Alignment ​

Author: Yi Nian, Tiankai Yang, Yudi Zhang, Qi Pan, Zelong Xu, Shenzhe Zhu, Qingqing Luan, Yue Huang, Xiangliang Zhang, Yue Zhao
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.07678v4 Announce Type: replace Abstract: Safety alignment for large language models relies on preference data, but current pipelines often train on large, redundant datasets. Existing data selection methods typically score each preference pair independently, collapsing directional prefere...

📖 Read original article


195. WhiFlash: Accelerating Speculative Decoding with Token-Level Cross-Paradigm Routing ​

Author: Young D. Kwon, Miles Williams, Rui Li, Alexandros Kouris, Stylianos I. Venieris
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.07710v2 Announce Type: replace Abstract: The autoregressive nature of large language models (LLMs) remains a significant bottleneck for inference, particularly in complex agentic workloads. While speculative decoding (SD) accelerates inference, current approaches rely on static drafting p...

📖 Read original article


196. A Geometry-Aware Triplane Field Network for Vehicle Aerodynamic Prediction ​

Author: Kangkang Qi, Huiyu Yang, Keqi Ding, Yunpeng Wang, Yuntian Chen, Yuanwei Bin, Rikui Zhang, Jianchun Wang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.07724v2 Announce Type: replace Abstract: High-fidelity computational fluid dynamics (CFD) is crucial to vehicle aerodynamic analysis, but its cost still constrains early-stage design exploration. Machine-learning-based surface-field prediction offers a faster alternative if the model can ...

📖 Read original article


197. Emotional regulation improves deep learning-based image classification ​

Author: Riccardo Emanuele Landi, Jo~ao M. F. Rodrigues, Marta Chinnici
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.13081v2 Announce Type: replace Abstract: Emotion significantly influences cognition, enhancing memory and learning under certain conditions. Drawing on this principle, emotion-augmented deep learning investigates how affective states can improve neural network architectures and learning p...

📖 Read original article


198. MM++: Post-Hoc Scale-Invariant Multilayer OOD Detection via Top-K Gated Feature Fusion ​

Author: Rahim Hossain, Md Tawheedul Islam Bhuian, Md Farhan Shadiq, Kyoung-Don Kang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2606.17352v2 Announce Type: replace Abstract: We introduce MM++ (Multilayer Mahalanobis++), a strictly post-hoc, and scale-invariant framework for Out-of-Distribution (OOD) detection. To address the trade-off between scale invariance and hierarchical expressivity, MM++ constructs a principled ...

📖 Read original article


199. Objective-Behavior Alignment: Diagnostics for MORL Policy Selection ​

Author: Antonio Mone, Zuzanna Osika, Florian Felten, Pradeep K. Murukannaiah, Mark Fuge, Frans A. Oliehoek, Luciano Cavalcante Siebert
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.21321v2 Announce Type: replace Abstract: Real-world decision-making often requires optimizing multiple competing objectives simultaneously. In reinforcement learning (RL), this is typically addressed by combining reward signals into a single scalar objective via a scalarization function, ...

📖 Read original article


200. TaLK: Text-attributed Graph Dataset Distillation via Coupling Language Model with Graph-Aware Kernel ​

Author: Yeongho Kim, Yeonje Choi, Kijung Shin
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.22975v3 Announce Type: replace Abstract: Text-attributed graphs (TAGs) are widely used in many real-world domains, and learning on TAGs requires jointly modeling text semantics and graph structure. A standard approach for modeling TAGs is to combine a language model (LM) and a graph neura...

📖 Read original article


201. AdaBoosting Text Prompts for Vision-Language Models ​

Author: Seokhee Jin, Changhwan Sung, Sunung Mun, Hoyoung Kim, Jungseul Ok
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.00684v4 Announce Type: replace Abstract: The classification accuracy of pretrained Vision-Language Models (VLMs) relies on the quality of the text prompts. Handcrafted templates and Large Language Model (LLM)-generated descriptions not only make predictions more interpretable, but also en...

📖 Read original article


202. Gauge dependence and structured-output corruption in sign-branched repetition penalties: measurements across models, inference stacks, and alternative repetition controls ​

Author: Peter Hollows
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.09791v2 Announce Type: replace Abstract: The multiplicative repetition penalty shipped across the LLM inference ecosystem (HuggingFace, vLLM, llama$.$cpp, and a dozen further engines) branches on the sign of each raw logit (divide positives by theta, multiply negatives). But the softmax i...

📖 Read original article


203. The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Shared Category Geometry in Small Language Models ​

Author: Francesco Karim Vicidomini
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.16741v3 Announce Type: replace Abstract: B"urger et al.\ (2024) demonstrated that truth representations in large language models are universal across statement polarity but reside within a multidimensional subspace. The truth value of a statement is linearly readable from a residual stre...

📖 Read original article


204. Persistent Sparse Autoencoders: Learning Feature-Specific Timescales in Language Model Representations ​

Author: Haoyan Luo, Mateo Espinosa Zarlenga, Mateja Jamnik
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.17117v2 Announce Type: replace Abstract: Sparse autoencoders (SAEs) decompose language model activations into sparse features, yet these models traditionally encode each token independently, failing to expose information that persists across a sequence. We first show that temporal persist...

📖 Read original article


205. One Model, Many Graphs: Learning over Attributed Graphs across Heterogeneous Modalities with Vision-Language Models ​

Author: Jiayi Yang, Yifang Chen, Yuanfu Sun, Jiajin Liu, Qiaoyu Tan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19128v2 Announce Type: replace Abstract: Vision-language models (VLMs) provide a unified representation space for textual and visual information, yet their potential as general-purpose backbones for graph-structured data remains largely unexplored. In practice, attributed graphs exhibit s...

📖 Read original article


206. Held-out evidence resolves follow-up measurement decisions in biological screens ​

Author: Jia Bi, Samuel Pinilla, Chenyang Zhu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.27651v3 Announce Type: replace Abstract: Machine learning determines which follow-up measurements biological screens collect. In a six-rule Cell Painting battery, the highest-value rule would re-image 96.01% of the library and had a 97.14% false-activation upper bound, showing why predict...

📖 Read original article


207. Feature Interaction Modeling for Neural Operators ​

Author: Quan Gu, Xiaoduo Li, Hongxia Liu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.28762v2 Announce Type: replace Abstract: Despite the many variants of DeepONet that have been proposed, query-based operator networks still struggle with shock-dominated and low-viscosity PDEs, whose sharp moving discontinuities and slowly decaying solution spectra challenge finite-dimens...

📖 Read original article


208. Training nGPT ​

Author: Ilya Loshchilov, Boris Ginsburg
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.01284v2 Announce Type: replace Abstract: The normalized Transformer (nGPT) realizes hyperspherical representation learning by constraining model parameter vectors and activation vectors to the unit hypersphere. In this paper, we describe a practical training recipe for nGPT and evaluate i...

📖 Read original article


209. Scaling an Autoregressive Transformer for Single-Cell Generation ​

Author: Aleksandr Sharipov, Yusif Mukhtarov, Igor Molybog
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.GN

arXiv:2608.02961v2 Announce Type: replace Abstract: We study a self-supervised generation task for single-cell gene expression vectors: given a set of vectors from a cell type, we aim to generate additional gene expression vectors of that cell type. For this task we characterize both the biological ...

📖 Read original article


210. How Far Do Simple Transformations Translate Across Text Embedding Models? ​

Author: Sid Ali Hamideche (Orange Research), Louis-Adrien Dufr`ene (Orange Research), Quentin Lampin (Orange Research), Guillaume Larue (Orange Research)
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.05980v2 Announce Type: replace Abstract: We investigate whether simple transformations can translate representations across heterogeneous text embedding models. Understanding how independently trained models organize semantic information is an enabler for AI-to-AI latent communication wit...

📖 Read original article


211. TransfHAR: Self-Supervised Wrist Representations for On-Demand Activity Recognition ​

Author: Aidan Bradshaw, Riku Arakawa, Xin Liu, Karan Ahuja
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15861v2 Announce Type: replace Abstract: Fine-grained wrist activity recognition can support applications such as procedural step guidance and context-aware assistance, yet acquiring labeled data for every new task, user, and activity granularity remains a bottleneck. We present TransfHAR...

📖 Read original article


212. Mol-JEPA: A multimodal Joint Embedding Predictive Architecture for Molecules ​

Author: Florian Rottach, Sebastian Schieferdecker, William Rudman, Randall Balestriero, Carsten Eickhoff
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22642v4 Announce Type: replace Abstract: Despite recent advances in molecular foundation models, several limitations remain, such as chemically invalid augmentations, modality collapse, and incomplete representation of biochemical environments. To address these challenges, we present \tex...

📖 Read original article


213. The Axiomatic Trader: Latent Regularity, Information Budgets, and the Canonical Form of a Quantitative Investment System ​

Author: Jiayu Li
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, q-fin.PM

arXiv:2608.23416v2 Announce Type: replace Abstract: Systematic trading rests on one article of faith: that regularities found in the past persist. This paper does three things. First, it states that faith as five axioms, each a commonplace practitioners already accept: (A1) a decision may use only w...

📖 Read original article


214. A Feature-Major Codebook for Memory-Efficient Sparse-Binary Self-Organizing Maps: Scaling a MEDLINE Atlas to 1.05 Million Neurons on a Single Consumer GPU ​

Author: Andrew James Amos
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2608.24067v2 Announce Type: replace Abstract: Building a self-organising map at MEDLINE scale has been impractical: the best-matching-unit (BMU) search that dominates training is bound by the bandwidth needed to read the codebook every epoch. I show that this bottleneck is largely an artefact ...

📖 Read original article


215. A Storage-Retrieval Gap in Parametric Knowledge Graph Memory ​

Author: Martino M. L. Pulici, Cuong Xuan Chu, Evgeny Kharlamov, Volker Tresp
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IR

arXiv:2608.25489v2 Announce Type: replace Abstract: Graph retrieval-augmented generation places retrieved subgraphs into the model's context window at query time, paying a recurring token cost and exposing source data on every call. We study an alternative: compiling a knowledge graph offline into a...

📖 Read original article


216. Tracing Generated Samples to Training-Data Clusters in Flow-Matching Models ​

Author: Rania Briq, Ohad Fried, Michael Kamp, Stefan Kesselheim
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.30081v2 Announce Type: replace Abstract: Understanding which training samples influence a generated image is an important problem in generative modeling. In flow matching, training samples influence the generated image through the velocity field along the generation trajectory. Removing s...

📖 Read original article


217. QTEA: Ternary LLMs with Sparse Residual Salient Weight and By-Column Optimization ​

Author: Yipin Guo, Arun M George, Jie Fu, Tareq Mahmoud, Sixue Xing, Siddharth Joshi
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.00224v2 Announce Type: replace Abstract: Weight-only post-training quantization (PTQ) can alleviate the computational burden of serving large language models (LLMs) at scale. However, existing PTQ methods often fail to generalize across models and suffer severe accuracy loss below 2 bits....

📖 Read original article


218. Why Multi-Layer Message Passing Works: Completeness Theory for Graph Neural Network Interatomic Potentials ​

Author: Pingbing Ming, Han Wang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, physics.chem-ph, physics.comp-ph

arXiv:2609.00528v2 Announce Type: replace Abstract: We prove that the Hypergraph Neural Network, an invariant architecture with 3-body message passing, is a universal approximator for potential energy surfaces. Our main contribution is a multi-layer completeness theory. We show that $L$ layers of me...

📖 Read original article


219. EEG-VID: Task-Guided Latent Predictive Pretraining for EEG Decoding and Assistive Target Selection ​

Author: Guanzhong Sun, Junyi Ma, Yuxuan Wu, Yanzi Miao
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.00566v2 Announce Type: replace Abstract: We propose EEG-VID, a task-guided latent predictive pretraining framework for EEG decoding under session and subject shifts. EEG-VID predicts future latent EEG states from recent history using an exponential-moving-average target encoder and weak t...

📖 Read original article


220. Let Confidence Change, Not the Prediction: Prediction-Preserving Repair for Post-hoc Calibration ​

Author: Daehwan Kim, Haejun Chung, Ikbeom Jang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2609.01072v2 Announce Type: replace Abstract: Post-hoc calibration corrects reported confidence, yet a multiclass calibrator can also change the associated top-1 prediction. Accuracy captures only the net effect of these changes on correctness, not how often predictions change; the Top-1 Predi...

📖 Read original article


221. Subliminal Learning as Trait-Direction Drift: A Mechanism and Targeted Control under SFT Distillation ​

Author: Zhixuan Liu, Zhichen Dong, Yuyu Fan, Xiangtian Li, Chao Yang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2609.01091v2 Announce Type: replace Abstract: Beyond intended capabilities, model distillation can transfer hidden traits from a teacher. A teacher biased by a system prompt can generate semantically clean training data, such as numeric sequences, that still causes a downstream student to inhe...

📖 Read original article


222. Bandits in Prod: Hyperparameter Optimization at Inference Time ​

Author: Louis Abraham, Tuan-Anh Nguyen, Nicolas Devatine
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2609.01335v2 Announce Type: replace Abstract: Many production systems can assess a configuration only by using it on live requests and observing noisy feedback. Modern agentic systems are a prominent example, with inference-time choices such as model selection, retrieval depth, prompting strat...

📖 Read original article


223. Rethinking Learnability in Offline Data-driven Optimization ​

Author: Chao Qian, Chen-Guang Wang, Rong-Xi Tan, Ke Xue
Published: 9/3/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE

arXiv:2609.01493v2 Announce Type: replace Abstract: Black-Box Optimization (BBO) has broad applications, while traditional algorithms such as evolutionary algorithms and Bayesian optimization face efficiency challenges as real-world BBO problems grow increasingly complex. Data-driven optimization ha...

📖 Read original article


224. Robust Streaming PCA ​

Author: Daniel Bienstock, Minchan Jeong, Apurv Shukla, Se-Young Yun
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:1902.03223v4 Announce Type: replace-cross Abstract: We consider streaming principal component analysis when the stochastic data generating model is subject to perturbations. While existing models assume a fixed covariance, we adopt a robust perspective where the covariance matrix belongs to a ...

📖 Read original article


225. Generalized Regret Analysis of Thompson Sampling using Fractional Posteriors ​

Author: Prateek Jaiswal, Debdeep Pati, Anirban Bhattacharya, Bani K. Mallick
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.SY, eess.SY, math.OC, math.ST, stat.TH

arXiv:2309.06349v2 Announce Type: replace-cross Abstract: Thompson sampling (TS) is one of the most popular and earliest algorithms to solve stochastic multi-armed bandit problems. We consider a variant of TS, named $\alpha$-TS, where we use a fractional or $\alpha$-posterior ($\alpha\in(0,1)$) inst...

📖 Read original article


226. Clustering Three-Way Data with Outliers ​

Author: Katharine M. Clark, Paul D. McNicholas
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2310.05288v4 Announce Type: replace-cross Abstract: Matrix-variate distributions are a relatively recent addition to the model-based clustering literature, thereby making it possible to analyze data in matrix form with complex structure such as images and time series. Due to its recent appeara...

📖 Read original article


227. GPTBIAS: A Comprehensive Framework for Evaluating Bias in Large Language Models ​

Author: Jiaxu Zhao, Meng Fang, Shirui Pan, Wenpeng Yin, Mykola Pechenizkiy
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.CY, cs.LG

arXiv:2312.06315v2 Announce Type: replace-cross Abstract: Warning: This paper contains content that may be offensive or upsetting. There has been a significant increase in the usage of large language models (LLMs) in various applications, both in their original form and through fine-tuned adaptation...

📖 Read original article


228. Deep Reinforcement Learning for Reach-Avoid-Stay Problems ​

Author: Gabriel Chenevert, Jingqi Li, Achyuta kannan, Sangjae Bae, Donggun Lee
Published: 9/3/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.RO, cs.SY

arXiv:2410.02898v3 Announce Type: replace-cross Abstract: Reach-Avoid-Stay (RAS) tasks are essential in applications where systems must safely reach a target set and remain within it under all bounded disturbances. Existing approaches either struggle to compute the maximal robust RAS set, the set of...

📖 Read original article


229. Enhancing brain age estimation with structural MRI and synthesized cerebral blood volume maps ​

Author: Jordan Jomsky, Zongyu Li, Kay C. Igwe, Yiren Zhang, Max Lashley, Tal Nuriel, Andrew Laine, Scott A. Small, Jia Guo, for the Frontotemporal Lobar Degeneration Neuroimaging Initiative, for the Alzheimer's Disease Neuroimaging Initiative
Published: 9/3/2026, 4:00:00 AM
Categories: eess.IV, cs.LG

arXiv:2412.01865v5 Announce Type: replace-cross Abstract: BrainAGE is a promising imaging-derived biomarker of neurobiological ageing and disease risk, yet current approaches rely predominantly on T1-weighted structural MRI, overlooking functional vascular changes that may precede tissue damage and ...

📖 Read original article


230. Sample Complexity of Linear Quadratic Regulator Without Initial Stability ​

Author: Amirreza Neshaei Moghaddam, Alex Olshevsky, Bahman Gharesifard
Published: 9/3/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.SY, eess.SY

arXiv:2502.14210v4 Announce Type: replace-cross Abstract: Inspired by REINFORCE, we introduce a novel receding-horizon algorithm for the Linear Quadratic Regulator (LQR) problem with unknown dynamics. Unlike prior methods, our algorithm avoids reliance on two-point gradient estimates while maintaini...

📖 Read original article


231. Quantum Speedups for Sampling and Non-convex Optimization with Stochastic Oracles ​

Author: Guneykan Ozgul, Xiantao Li, Mehrdad Mahdavi, Chunhao Wang
Published: 9/3/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, math.OC

arXiv:2504.03626v2 Announce Type: replace-cross Abstract: We present quantum speedups for sampling from distributions of the form $\pi\propto e^{-f}$ on $\mathbb{R}^d$. We consider two stochastic oracle models: a stochastic gradient oracle, where $f=\frac{1}{n}\sum_{i=1}^n f_i $ and component gradie...

📖 Read original article


232. DLM-One: Diffusion Language Models for One-Step Sequence Generation ​

Author: Tianqi Chen, Shujian Zhang, Mingyuan Zhou
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, stat.ML

arXiv:2506.00290v2 Announce Type: replace-cross Abstract: This paper introduces DLM-One, a score-distillation-based framework for one-step sequence generation with continuous diffusion language models (DLMs). DLM-One eliminates iterative refinement by aligning the scores of a student model's outputs...

📖 Read original article


233. Learning Encodings by Maximizing State Distinguishability: Variational Quantum Error Correction ​

Author: Nico Meyer, Christopher Mutschler, Andreas Maier, Daniel D. Scherer
Published: 9/3/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2506.11552v3 Announce Type: replace-cross Abstract: Quantum error correction is crucial for protecting quantum information against decoherence. Traditional codes like the surface code require substantial overhead, making them impractical for near-term, early fault-tolerant devices. We propose ...

📖 Read original article


234. Explainable Information Processing in Particle Swarm Optimization through Landscape and Search Behavior Analysis ​

Author: Nitin Gupta, Bapi Dutta, Anupam Yadav
Published: 9/3/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2509.06272v5 Announce Type: replace-cross Abstract: Swarm-based optimization algorithms have demonstrated remarkable success in solving complex problems, yet their widespread adoption remains limited due to poor transparency in how algorithmic components influence performance. This work presen...

📖 Read original article


235. Adversarial Stress Testing of Outlier Detection in Subjective Image Quality Assessment ​

Author: Dietmar Saupe
Published: 9/3/2026, 4:00:00 AM
Categories: eess.IV, cs.LG, cs.MM

arXiv:2509.06554v2 Announce Type: replace-cross Abstract: In subjective image and video quality assessment, observers rate or compare selected stimuli. Before calculating mean opinion scores (MOSs), unreliable ratings should be identified and handled as outliers. Several outlier-detection methods ar...

📖 Read original article


236. Toward Uncertainty-Aware and Generalizable Neural Decoding for Quantum LDPC Codes ​

Author: Xiangjun Mi, Frank Mueller
Published: 9/3/2026, 4:00:00 AM
Categories: quant-ph, cs.IT, cs.LG, math.IT

arXiv:2510.06257v2 Announce Type: replace-cross Abstract: Quantum error correction (QEC) is essential for scalable quantum computing, yet decoding errors via conventional algorithms result in limited accuracy (i.e., suppression of logical errors) and high overheads, both of which can be alleviated b...

📖 Read original article


237. Neural Variational Cut Posteriors without Upstream Data ​

Author: Jiafang Song, Sandipan Pramanik, Abhirup Datta
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2510.10268v3 Announce Type: replace-cross Abstract: In many applications, one must propagate parameter uncertainty from an earlier (upstream) analysis, available as samples, to subsequent (downstream) analyses without feedback. This problem is called cutting feedback or cut-Bayes, and the cut-...

📖 Read original article


238. GMTRouter: Personalized LLM Router over Multi-turn User Interactions ​

Author: Yihang Sun, Encheng Xie, Tao Feng, Jiaxuan You
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2511.08590v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) routing has demonstrated strong capability in balancing response quality with computational cost. As users exhibit diverse preferences, personalization has attracted increasing attention in LLM routing, since even i...

📖 Read original article


239. Enhancing Road Safety Through Multi-Camera Image Segmentation with Post-Encroachment Time Analysis ​

Author: Shounak Ray Chaudhuri, Arash Jahangiri, Christopher Paolini
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.SI

arXiv:2511.12018v2 Announce Type: replace-cross Abstract: Traffic safety analysis at signalized intersections is essential for reducing vehicle and pedestrian collisions, yet traditional crash-based studies are limited by data sparsity and reporting latency. This paper presents a multi-camera comput...

📖 Read original article


240. Secure AI-Driven Super-Resolution for Real-Time Mixed Reality Applications ​

Author: Mohammad Waquas Usmani, Sankalpa Timilsina, Michael Zink, Susmit Shannigrahi
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.MM, eess.IV

arXiv:2512.15823v3 Announce Type: replace-cross Abstract: Immersive formats such as 360{\deg} and 6DoF point cloud videos require high bandwidth and low latency, posing challenges for real-time AR/VR streaming. This work focuses on reducing bandwidth consumption and encryption/decryption delay, two ...

📖 Read original article


241. On Cost-Aware Designs for Sequential Hypothesis Testing ​

Author: George Vershinin, Asaf Cohen, Omer Gurewitz
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT

arXiv:2512.19067v2 Announce Type: replace-cross Abstract: We introduce Cost-Aware (CA) Sequential Hypothesis Testing (CASHT), in which an active decision-maker selects sensing actions with differing, random costs to identify the true hypothesis under an average-error constraint $\delta$ while minimi...

📖 Read original article


242. What Drives Success in Physical Planning with Joint-Embedding Predictive World Models? ​

Author: Basile Terver, Tsung-Yen Yang, Jean Ponce, Adrien Bardes, Yann LeCun
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.RO, stat.ML

arXiv:2512.24497v4 Announce Type: replace-cross Abstract: A long-standing challenge in AI is to develop agents capable of solving a wide range of physical tasks and generalizing to new, unseen tasks and environments. A popular recent approach involves training a world model from state-action traject...

📖 Read original article


243. Beyond Transfer Accuracy: Mechanism-Guided Controlled Adaptation for Low-Resource Languages ​

Author: Khumaisa Nur'aini, Ayu Purwarianti, Alham Fikri Aji, Derry Wijaya
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2601.08146v4 Announce Type: replace-cross Abstract: Existing circuit discovery methods rely on templated tasks with clean counterfactuals, limiting their use on diverse natural text. We adapt Contextual Decomposition for Transformers (CD-T) for unstructured settings via label-balanced activati...

📖 Read original article


244. Learning and extrapolating scale-invariant processes ​

Author: Anaclara Alvez-Canepa, Cyril Furtlehner, Fran\c{c}ois P. Landes
Published: 9/3/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cs.LG

arXiv:2601.14810v3 Announce Type: replace-cross Abstract: Machine Learning (ML) has deeply changed some fields recently, like Language and Vision and we may expect it to be relevant also to the analysis of of complex systems. Here we want to tackle the question of how and to which extent can one reg...

📖 Read original article


245. Towards Solving the Gilbert-Pollak Conjecture via Large Language Models ​

Author: Yisi Ke, Tianyu Huang, Yankai Shu, Di He, Jingchu Gai, Liwei Wang
Published: 9/3/2026, 4:00:00 AM
Categories: cs.DM, cs.LG

arXiv:2601.22365v3 Announce Type: replace-cross Abstract: The Gilbert-Pollak Conjecture \citep{gilbert1968steiner}, also known as the Steiner Ratio Conjecture, states that for any finite point set in the Euclidean plane, the Steiner minimum tree has length at least $\sqrt{3}/2 \approx 0.866$ times t...

📖 Read original article


246. Modular Expert Merging for Biomedical Retrieval ​

Author: Sameh Khattab, Jean-Philippe Corbeil, Osman Alperen \c{C}inar-Kora\c{s}, Amin Dada, Julian Friedrich, Jiawei He, Douglas Teodoro, Jens Kleesiek
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2602.04731v3 Announce Type: replace-cross Abstract: Adapting general-purpose LLMs into domain-specialized dense retrievers typically requires large-scale training on mixed-domain data. We show that merging independently trained domain-specialized experts consistently exceeds this approach acro...

📖 Read original article


247. Quantum Maximum Likelihood Prediction via Hilbert Space Embeddings ​

Author: Sreejith Sreekumar, Nir Weinberger
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT, quant-ph, stat.ML

arXiv:2602.18364v4 Announce Type: replace-cross Abstract: Maximum likelihood prediction (MLP) is a core task at the heart of modern large language models. Here, we study a quantum version of this task for a simplified data model consisting of independent and identically distributed samples, as a fir...

📖 Read original article


248. DynaTokens: Controlling Token Dynamics for Continual Video-Language Understanding ​

Author: Toan Nguyen, Yang Liu, Celso De Melo, Flora D. Salim
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2603.06662v3 Announce Type: replace-cross Abstract: Continual VideoQA with multimodal LLMs remains challenging because sequential adaptation induces task interference, while storing task-specific prompts becomes impractical as task sequences grow. We introduce DynaTokens, a transformer-based t...

📖 Read original article


249. GONE: Structural Knowledge Unlearning via Neighborhood-Expanded Distribution Shaping ​

Author: Chahana Dahal, Ashutosh Balasubramaniam, Zuobin Xiong
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2603.12275v2 Announce Type: replace-cross Abstract: Unlearning knowledge is a pressing and challenging task in Large Language Models (LLMs) because of their unprecedented capability to memorize and digest training data at scale, raising more significant issues regarding safety, privacy, and in...

📖 Read original article


250. Probing Cultural Signals in Large Language Models through Author Profiling ​

Author: Valentin Lafargue, Ariel Guerra-Adames, Emmanuelle Claeys, Elouan Vuichard, Jean-Michel Loubes
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2603.16749v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed in applications with societal impact, raising concerns about the cultural biases they encode. We probe these representations by evaluating whether LLMs can perform author profiling from s...

📖 Read original article


251. ICE: Intervention-Consistent Explanation Evaluation with Statistical Grounding for LLMs ​

Author: Abhinaba Basu, Pavan Chakraborty
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2603.18579v2 Announce Type: replace-cross Abstract: Evaluating whether explanations faithfully reflect a model's reasoning remains an open problem. Existing benchmarks use single interventions without statistical testing, making it impossible to distinguish genuine faithfulness from chance-lev...

📖 Read original article


252. From High-Dimensional Spaces to Verifiable ODD Coverage for Safety-Critical AI-based Systems ​

Author: Thomas Stefani, Johann Maximilian Christensen, Elena Hoemann, Frank K"oster, Sven Hallerbach
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2604.02198v2 Announce Type: replace-cross Abstract: While Artificial Intelligence (AI) offers transformative potential for operational performance, its deployment in safety-critical domains such as aviation requires strict adherence to rigorous certification standards. Current EASA guidelines ...

📖 Read original article


253. Conditional Diffusion Posterior Alignment for Sparse-View CT Reconstruction ​

Author: Luis Barba, Johannes Kirschner, Benjamin Bejar
Published: 9/3/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2604.21960v3 Announce Type: replace-cross Abstract: Computed Tomography (CT) is a widely used imaging modality in medical and industrial applications. To limit radiation exposure and measurement time, there is a growing interest in sparse-view CT, where the number of projection views is signif...

📖 Read original article


254. Stabilizing Private LASSO under Heterogeneous Covariates via Anisotropic Objective Perturbation ​

Author: Haruka Tanzawa, Ayaka Sakata
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT

arXiv:2605.01492v2 Announce Type: replace-cross Abstract: We study high-dimensional LASSO under differential privacy via objective perturbation with heterogeneous covariate scales. In practical scenarios, covariates often exhibit diverse scales; however, standard preprocessing is problematic under p...

📖 Read original article


255. Connections between the F\"ollmer process and the denoising diffusion probabilistic model ​

Author: Yuta Koike
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR

arXiv:2605.18040v2 Announce Type: replace-cross Abstract: The F"ollmer process is a Brownian motion conditioned to have a pre-specified distribution at time 1. This process can be interpreted as an ``augmented'' time-compressed version of the reverse stochastic differential equation (SDE) correspon...

📖 Read original article


256. Half-Truth Audio Detection and Localisation: A Lightweight Cross-Attentive Architecture and a Cross-Corpus Diagnostic Study ​

Author: S. Sutharya, Remya K. Sasi
Published: 9/3/2026, 4:00:00 AM
Categories: cs.SD, cs.CV, cs.LG

arXiv:2605.29531v3 Announce Type: replace-cross Abstract: Partially manipulated (half-truth) speech, where a short synthesised segment is spliced into an otherwise genuine utterance, is a harder and more realistic forensic threat than the fully synthesised deepfakes that dominate the literature. We ...

📖 Read original article


257. OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation ​

Author: Haochen Yang, Ke Zhao, Mengyuan Ma, Xingyu Lu, Xiangfeng Wang, Hong Qian
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2605.29829v2 Announce Type: replace-cross Abstract: Leveraging Large Language Models (LLMs) to automatically formulate and solve optimization problems from natural language has emerged as an efficient paradigm for automated optimization. However, existing methods still exhibit limited generali...

📖 Read original article


258. Variation Spaces for Encoder--Decoder Neural Operators: Approximation and Generalization ​

Author: Jia-Qi Yang, Lei Shi
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.FA, math.NA, math.ST, stat.TH

arXiv:2606.01244v2 Announce Type: replace-cross Abstract: Inspired by the function-space theory of neural networks, we formulate and analyze a variation space for nonlinear operators between Hilbert spaces, defined through vector-valued Borel measures of bounded variation. We characterize its unit b...

📖 Read original article


259. Medical Heuristic Learning: An LLM-Driven Framework for Interpretable and Auditable Clinical Decision Rules ​

Author: Wei Xu, Ke Yang, Gang Luo, Keli Zheng, Lingyan Hu, Jing Wang, Kefeng Li
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.HC, cs.LG

arXiv:2606.16337v4 Announce Type: replace-cross Abstract: Predictive modeling for clinical decision support requires both strong predictive performance and transparent, auditable, and human-reviewable decision logic. Although deep learning and tree-based ensemble methods can achieve high accuracy, t...

📖 Read original article


260. SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics ​

Author: Nikolay Georgiev, Maria Drencheva, Kseniia Ibragimova, Ivo Petrov, Dimitar I. Dimitrov, Martin Vechev
Published: 9/3/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG

arXiv:2606.29894v2 Announce Type: replace-cross Abstract: As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theorem libraries, and educational resources. However, choosing the right retriever remains diffic...

📖 Read original article


261. What You See Is What You Get: Observation-Aligned Supervision for Chart-to-Code Generation ​

Author: Tianhao Niu, Qingfu Zhu, Wanxiang Che
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.04726v5 Announce Type: replace-cross Abstract: Chart-to-code generation is commonly trained through supervised fine-tuning on reference plotting scripts, implicitly treating the gold code as a fully observable target. However, many chart programs contain latent variables that cannot be un...

📖 Read original article


262. Multi-Mask Diffusion Language Models for Few-Step Generation ​

Author: Sijin Chen, Yinuo Ren, Heyang Zhao, Ziheng Cheng, Quanquan Gu, Lexing Ying
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.19686v3 Announce Type: replace-cross Abstract: Masked diffusion models (MDMs) are a promising family of language generators, but achieving high-quality few-step generation remains challenging. In MDMs, all forward trajectories collapse to a single fully masked state, leaving no terminal e...

📖 Read original article


263. Reinforcement Learning for Heterogeneous Sensor Selection in Maritime Surveillance ​

Author: Andrei Starodubov, Yaqub Aris Prabowo, Andreas Hadjipieris, Roberto Galeazzi, Ioannis Kyriakides
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.IT, cs.LG, cs.RO, cs.SY, eess.SP, eess.SY, math.IT

arXiv:2607.22667v2 Announce Type: replace-cross Abstract: This paper presents an information-gain-guided reinforcement-learning sensor-selection framework for single-vessel tracking in heterogeneous maritime sensor networks. The proposed approach is motivated by information-theoretic sensor manageme...

📖 Read original article


264. Adaptive Graph-of-Islands Evolution for Automatic Feature Engineering with LLMs ​

Author: Sha Li, Naren Ramakrishnan
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.23286v2 Announce Type: replace-cross Abstract: Automatic feature engineering (AutoFE) for tabular data requires discovering informative transformations from a large program space. Existing approaches suffer from three limitations: classical methods rely on fixed operator libraries with li...

📖 Read original article


265. Aletheia: An Offline-First Clinical Decision Support System for Differential Diagnosis in Low-Resource Healthcare Settings ​

Author: Joseph Walusimbi, Ann Move Oguti, Abubakhari Sserwadda, Precious Boss Kasasira, Charles Brian Okoboi
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-bio.OT

arXiv:2607.24814v2 Announce Type: replace-cross Abstract: Access to specialist clinical expertise remains severely limited across sub-Saharan Africa, where physician-to-patient ratios can fall below 1:25,000 in rural settings. Existing AI-assisted diagnostic tools predominantly require reliable inte...

📖 Read original article


266. Windowed thinning and query complexity for the bouncy particle and Zigzag samplers ​

Author: Jianfeng Lu, Yinchen Luo
Published: 9/3/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.PR, math.ST, stat.CO, stat.TH

arXiv:2607.28413v2 Announce Type: replace-cross Abstract: Let $\mu(d x)\propto e^{-U(x)} d x$ on $\R^d$, where $U$ is $m$-strongly convex and $L$-smooth, and denote by $\kappa=L/m$ the condition number. We consider windowed thinning, an exact simulation method for the bouncy particle sampler and the...

📖 Read original article


267. Nova: An End-to-End MLIR Compiler for Deep Learning ​

Author: Adwaid Suresh, Aparna A, Harshini V M, Jona Delcy C A, Killi Uma Maheswara Rao, Ram Charan Golla, Surendra Vendra
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.AR, cs.LG, cs.PL

arXiv:2608.00029v3 Announce Type: replace-cross Abstract: The performance of deep learning models at scale relies heavily on how effectively high-level mathematical operations are mapped to underlying physical hardware. While high-level tensor frameworks provide flexible abstractions, their executio...

📖 Read original article


268. From Digital to Physical Reservoir Computing: Co-Optimizing Soft Robotic Reservoirs via Dynamics Matching ​

Author: Nicola Visentin, Maximilian St"olzle, Mariano Ram'irez Montero, Francesco Braghin, Daniela Rus, Cosimo Della Santina
Published: 9/3/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.00484v2 Announce Type: replace-cross Abstract: Soft robotic substrates are promising for Physical Reservoir Computing (PRC) because their compliant nonlinear dynamics can provide temporal memory, high-dimensional state transformations, and efficient inference. However, physical reservoirs...

📖 Read original article


269. Three Necessary Principles for Self-Supervised Visual Representation Learning ​

Author: Nikos Giakoumoglou, Paschalis Giakoumoglou, Tania Stathaki
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.08309v2 Announce Type: replace-cross Abstract: We argue that learning visual representations without labels requires a training signal jointly complete across three non-overlapping objectives: semantic invariance across augmented views, patch-level spatial prediction, and representational...

📖 Read original article


270. Diagonal Multi-omics Integration of Heterogeneous Datasets ​

Author: Maksim V. Kukushkin, Mikhail S. Arbatskiy, Dmitriy E. Balandin, Alexey V. Churov
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.FA

arXiv:2608.16968v2 Announce Type: replace-cross Abstract: In this paper, we consider methods for the diagonal multi-omics integration of heterogeneous datasets. Several approaches to the nature of biological heterogeneity are analyzed and developed to comprehend more clearly the generated difference...

📖 Read original article


271. FlavourBench: Executable Culinary Reward Maps for Language Model Evaluation and Post-Training ​

Author: Josef Chen (Independent Researcher), Erim Hayretci (Imperial College London)
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG, cs.SE

arXiv:2608.20574v3 Announce Type: replace-cross Abstract: We introduce FlavorBench: a benchmark for Compiling Dense Deterministic Answer Maps from a Versioned Culinary Embeddings Model. We test 27 frontier large language model endpoints on 534 substitution, pairing and constraining tasks for tasks t...

📖 Read original article


272. Agentic Scaffolding Amplifies Sycophantic Behavior in Large Language Models ​

Author: Thantham Jittham
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.MA

arXiv:2608.21377v2 Announce Type: replace-cross Abstract: Sycophancy in large language models, the tendency to prioritize user agreement over truthful responses, has been documented extensively but studied primarily in single-turn settings. This paper investigates a critical question: does subjectin...

📖 Read original article


273. ToSCA: Leveraging Hierarchical Reinforcement Learning on Temporal and Strategic Abstractions of Conversational Agents ​

Author: Xiaoyu Wang, Qingqing Gu, Yue Zhao, Teng Chen, Yuqi Cao, Xiaokai Chen, Hongyan Li, Luo Ji
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.HC, cs.LG

arXiv:2608.21969v3 Announce Type: replace-cross Abstract: Humans naturally exhibit multiple forms of abstraction in reasoning and interaction, including temporal abstraction across decision timescales and strategic abstraction over communicative intents. Inspired by these complementary abstractions,...

📖 Read original article


274. Scalable Self-Supervised Learning for Multiphase AC-OPF in Distribution Systems with Topology Reconfiguration ​

Author: Hoang T. Nguyen, Shaohui Liu, Reetam Sen Biswas, Varsha Pendyala, Nurali Virani, Deepjyoti Deka, Priya L. Donti
Published: 9/3/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC

arXiv:2608.25095v2 Announce Type: replace-cross Abstract: The proliferation of distributed energy resources (DERs) in distribution grids enables the active coordination of these assets to reduce costs and enable cleaner operations. Realizing this potential requires solving multiphase AC optimal powe...

📖 Read original article


275. Optimal Transport for Network Comparison: A Review with Machine Learning Applications ​

Author: James Hyun, Fran\c{c}ois G. Meyer
Published: 9/3/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.SI

arXiv:2608.27500v2 Announce Type: replace-cross Abstract: Network comparison using optimal transport is a growing area of research in network science. Unlike standard graph metrics, optimal transport computes both network dissimilarity and a transport plan that explains how one graph morphs into ano...

📖 Read original article


276. The Illusion of Replacement: Rethinking Specialized Machine Learning Models in the Foundation Model Era ​

Author: Kiyan Rezaee
Published: 9/3/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.28980v2 Announce Type: replace-cross Abstract: Can the specialized architectures that machine learning has traditionally built for structured data be replaced by language-based models? This question is examined through a review of 159 papers (2016--2026) across nine modalities, with predi...

📖 Read original article


277. VoiceLongMemEval: Do Assistants Remember How You Sounded? ​

Author: Ramit Pahwa, Parivesh Priye, Apoorva Beedu
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2609.00570v2 Announce Type: replace-cross Abstract: With the growing scale of multi-agent architectures and large language models, deployed AI assistants are increasingly tasked with reasoning over long, continuous, multi-session conversation histories. Current benchmarks evaluate this dialogu...

📖 Read original article


278. Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers ​

Author: Giovanni Bonetta, Matteo Merler, Davide Zago, Rossella Cancelliere, Bernardo Magnini
Published: 9/3/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2609.01567v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) provide useful priors for interactive decision-making, but using them directly as policies is expensive and brittle: they must be queried at every step, do not improve from environment interaction, and can repeat...

📖 Read original article