Skip to content

arXiv cs.LG - 2026-08-05 ​

302 items collected.


1. Deep Divide-and-Reduce in Symbolic Regression ​

Author: Yusong Deng, Yanjie Li, Weijun Li
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02628v1 Announce Type: new Abstract: Symbolic regression (SR) is the task of discovering underlying patterns from data and representing them using mathematical expressions. Current machine learning approaches to SR often lack a profound understanding of the intrinsic mathematical and phys...

📖 Read original article


2. Multimodal Auto-regressive Transformer Surrogate for Modeling Variable Operations and Quantifying Uncertainty in Geological Carbon Storage ​

Author: Yifu Han, Louis J. Durlofsky
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02629v1 Announce Type: new Abstract: The use of variable well perforation and injection strategies can improve the efficiency of geological carbon storage operations. We develop a new multimodal auto-regressive transformer surrogate to model these operations under geological uncertainty. ...

📖 Read original article


3. LLMs Can Annotate Attribution Graphs ​

Author: Ameen Patel, Max Zhang, Nathan Hu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02632v1 Announce Type: new Abstract: Circuit tracing is an exciting technique for revealing the internal computation of language models, but it requires a time-intensive manual step of grouping individual features or MLP neurons into supernodes. We present a simple pipeline for automating...

📖 Read original article


4. GeoID-PINN: Identifiability-Aware Regional Epidemic Inference with Geographic Coupling ​

Author: Weixiong Hua, Fan Bu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, stat.AP, stat.ML

arXiv:2608.02633v1 Announce Type: new Abstract: Regional surveillance data reflect local transmission, reporting, seeding, and external infection pressure, which are difficult to identify separately. We introduce GeoID-PINN, a physics-informed neural network (PINN) for susceptible-infectious-recover...

📖 Read original article


5. Verifier-Guided Model Discovery for Physical Dynamical Systems with Pretrained Symbolic Transformers ​

Author: Farbod Faraji, Francesco Belardinelli
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02662v1 Announce Type: new Abstract: Reliable forecasting of nonlinear physical systems underpins scientific discovery and engineering decision-making. Yet high-fidelity simulations are prohibitively costly, and machine-learning surrogates can be opaque and encode assumptions about system...

📖 Read original article


6. CT-HEG: A Bidirectional, Timestamp-Attributed Event Graph for ICU In-Hospital Mortality Prediction - An Architectural Ablation Study ​

Author: Mohammad Nasir Uddin, Rahnuma Tabassum Orpita, Asaduzzaman Anik, Eklachur Rahman Bhuiyan, Marjahan Risalat, SM Wali Ullah, Asif Ahamed
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02663v1 Announce Type: new Abstract: Accurate ICU mortality prediction requires modeling irregular clinical observations across heterogeneous entity types. Existing sequence models handle irregular sampling but ignore typed relational structure; existing graph models assume fixed-interval...

📖 Read original article


7. Sphere Retraction Normalizations ​

Author: Jie Zhang, Cheng-Fang Su, Yi-Jui Huang, Min-Te Sun
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.02668v1 Announce Type: new Abstract: Residual connections are the de facto mechanism for training deep neural networks stably. Geodesic Normalization (GeoNorm) recasts them on a Riemannian manifold, orthogonalizing each layer output against the current hidden state and applying the result...

📖 Read original article


8. Learning Molecular Representations from Cellular Phenotypes with Structure Preservation ​

Author: Xuan Lin, Jingyu Sheng, Tengfei Ma, Li Sun, Dapeng Xiong
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02688v1 Announce Type: new Abstract: Phenotypic drug discovery enables the discovery of functional relationships between molecular structures and cellular responses. However, existing multimodal representation learning methods often optimize cross-modal alignment without considering the i...

📖 Read original article


9. GLOBE: Trajectory-Aligned Gradient Matching with Structured SparseOptimization for Coreset Selection ​

Author: Hetian Liu, Jin Cui, Mengcheng Shi, Yanbin Hu, Xinyue Long, Boran Zhao, Pengju Pen
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02690v1 Announce Type: new Abstract: On-device training of deep neural networks is fundamentally constrained by the computational and memory costs of large-scale datasets. Coreset selection offers a practical solution by retaining only a compact subset of real training samples. However, e...

📖 Read original article


10. Output-Aware Rotation for INT2 KV-Cache Quantization ​

Author: Vincent-Daniel Yun, Woosang Lim, Minsoo Cheong, Sunwoo Lee, Murali Annavaram, Sai Praneeth Karimireddy, Sungjoo Yoo
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02691v1 Announce Type: new Abstract: The key-value (KV) cache has become a major memory and bandwidth bottleneck in long-context large language model inference, making ultra-low-bit quantization increasingly important. However, existing rotation-based INT2 methods optimize cache statistic...

📖 Read original article


11. PatTree: a novel approach for automated creation of multimodal, graph-based patient representations for medical classification tasks ​

Author: Julia Gehrmann, Lars Quakulinski, Hamza Naseem, Oya Beyan
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02692v1 Announce Type: new Abstract: Access to holistic, multimodal data improves the performance of Artificial Intelligence (AI) in medical classification tasks compared to utilizing single modalities or data sources. However, the inherent heterogeneity and complexity of clinical real-wo...

📖 Read original article


12. Measuring Explainer Stability via Attribution Separability ​

Author: Eddie Conti, 'Alvaro Parafita, Axel Brando
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02697v1 Announce Type: new Abstract: Attribution methods (AMs) assign an importance score to each feature and are widely adopted to explain black-box models. However, most methods can produce variable attribution scores due to stochastic components in their definition. In this paper, we p...

📖 Read original article


13. NANQ: Noise-Floor-Aware Mixed-Precision Non-Uniform Quantization for Analog Compute-in-Memory ​

Author: Yizhe Chen, Wenshuai Yao, Saiya Wang, Yuannuo Feng, Wenbo Qi, Kechao Tang, Ngai Wong, Wenyong Zhou, Wang Kang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02700v1 Announce Type: new Abstract: Analog compute-in-memory (CIM) enables energy-efficient neural network inference, but device variation and read noise can severely degrade low-bit quantized models. Existing CIM-oriented quantization methods mainly minimize ideal quantization error, ig...

📖 Read original article


14. Can Training Logs Make Model Comparisons More Precise? ​

Author: Wei-Jung Huang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.02705v1 Announce Type: new Abstract: Comparing stochastically trained models requires estimating both a performance difference and its uncertainty from repeated runs. We study whether training logs from those same runs can make such comparisons more precise. Because training-log covariate...

📖 Read original article


15. Designing a Good Virtual Node: Addressable and Cardinality-Preserving Global Memory for Message Passing Architectures ​

Author: F'elix Marcoccia
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02709v1 Announce Type: new Abstract: Virtual nodes give message-passing neural networks a simple global communication route, but the standard node--VN--node pipeline compresses the graph into one homogeneous state and broadcasts it identically to every node. Building on the Two-Radius ana...

📖 Read original article


16. Neural Networks with Local Converging Inputs for Efficient Options Pricing Models ​

Author: Harris Cobb, Wenbo Hao, Yingjie Liu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP

arXiv:2608.02778v1 Announce Type: new Abstract: We present a novel application of Neural Networks with Local Converging Inputs (NNLCI) to improve the efficiency of existing numerical methods for pricing multi-asset options. The most concise input format for NNLCI has been introduced, offering substa...

📖 Read original article


17. Evaluation Blindness: How Silent Measurement Failures Corrupt AI Systems from Training to Deployment ​

Author: Priyanka Bajaj (Independent Researcher)
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02786v1 Announce Type: new Abstract: AI systems can fail silently. The failure propagates through training loops, evaluation pipelines, and production monitoring stacks until downstream harm makes it visible. This paper introduces evaluation blindness: a measurement function M exhibits ev...

📖 Read original article


18. Topological Simplification in Predictive Coding Networks ​

Author: Adam Shaw, Jiayu Li, Michael Sperling, Michael Kim, Alvin Jin
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02816v1 Announce Type: new Abstract: We study the topology of learned representations in predictive coding networks (PCNs), a neuro-inspired bidirectional architecture, using a quantitative layer-wise persistent homology analysis. We train well-performing PCNs on a synthetic classificatio...

📖 Read original article


19. Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't ​

Author: Ravi Satya Durga Prasad Yenugula
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.02829v1 Announce Type: new Abstract: Model families train every size from scratch. Can a pretrained large model be converted into a smaller sibling? We characterize the 1.4B->410M conversion in the Pythia family end-to-end: (i) representations align strongly across sizes (ridge R^2=0.84) ...

📖 Read original article


20. NOMADD: Numerical Optimization of Models Adapting to Data Drift ​

Author: Swapn Shah, Keith Burghardt
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02845v1 Announce Type: new Abstract: Tabular model performance degrades when feature distributions change over time or the relationship between features and outcome variables change over time, known as data drift and concept drift, respectively. These issues are challenging to mitigate in...

📖 Read original article


21. Adaptive Sampling for Automated Post-Disaster Rapid Damage Assessment via Level-Set Cost-Aware Bayesian Optimization ​

Author: Boyang Xu, Mostafa Reisi Gahrooei, Mohammad Ilbeigi, Hao Yan
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02868v1 Announce Type: new Abstract: Natural disasters frequently inflict severe damage to the built environment, which demands a rapid, reliable, and cost-effective damage assessment for emergency response. However, traditional methods for post-disaster damage assessment often rely on st...

📖 Read original article


22. Contrast-invariant deep ptychography neural networks ​

Author: Albert Vong, Steven Henke, Oliver Hoidn, Hanna Ruth, Junjing Deng, Apurva Mehta, David Shapiro, Alexander Hexemer, Nicholas Schwarz
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02869v1 Announce Type: new Abstract: Ptychography neural networks suffer from scaling inconsistencies when generalizing out of distribution, limiting their real world viability. We address this scaling mismatch using a factorization strategy which decouples the learned object texture from...

📖 Read original article


23. Maglev: Sliding Recurrent Memory ​

Author: Bo Liu, Qiang Liu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02870v2 Announce Type: new Abstract: We introduce \ours{}, a recurrent Transformer architecture with fixed-size memory that generalizes sliding-window attention while remaining parallelizable during training. \ours{} consists of two coupled models: a prefiller $Q$, which leverages full at...

📖 Read original article


24. GoT-CD: Graph-of-Thoughts Causal Discovery and the Fragility of Post-hoc Path-Specific Fairness Audits ​

Author: Nitish Nagesh, Elahe Khatibi, Thomas Dean Hughes, Mahdi Bagheri, Pratik Gajane, Amir M. Rahmani
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02877v1 Announce Type: new Abstract: Causal discovery recovers directed structure from observational data and is increasingly used in clinical settings to support mechanism reasoning and fairness audits of predictive models. Path-specific counterfactual fairness asks whether a protected a...

📖 Read original article


25. Population-Robust Feature Selection via Generalized Welfare Optimization ​

Author: Ruiqi Lyu, Alistair Turcan, Bryan Wilder
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02887v1 Announce Type: new Abstract: Choosing which features to collect is a deployment decision: the same limited questionnaire, test panel, or sensor set may need to serve several heterogeneous populations. Standard feature-selection methods typically optimize for one large population, ...

📖 Read original article


26. Robust Counterfactual Policy Optimisation via Nondeterministic Causal Models ​

Author: Jessica Lally, Milad Kazemi, Nicola Paoletti, David Watson, Sander Beckers
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02893v1 Announce Type: new Abstract: Counterfactual inference approaches for sequential decision-making typically assume deterministic causal models, where all randomness stems from latent variables. However, Markov Decision Processes (MDPs) are inherently stochastic. We address this by f...

📖 Read original article


27. AnchorKV: Anchor-Residual KV Cache Compression ​

Author: Malik Khalaf, Yara Shamshoum, Nitzan Hodos, Yuval Sieradzki, Assaf Schuster
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.02901v1 Announce Type: new Abstract: The key-value (KV) cache is the primary memory bottleneck in long-context LLM inference. Existing approaches attack it from opposite ends: eviction methods permanently discard tokens, degrading performance whenever a discarded token later proves essent...

📖 Read original article


28. Bayesian Data Reweighting Improves Multimodal Retrieval for Knowledge-Based Visual Question Answering ​

Author: Jingchen Sun, Shaobo Han, Ruiyi Zhang, Naresh Kumar Devulapally, Ming Liu, Yitao Long, Vishnu Suresh Lokhande, Changyou Chen
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02907v1 Announce Type: new Abstract: Multimodal retrievers are essential for knowledge-based visual question answering, where they retrieve external evidence for image-question pairs. However, existing contrastive training methods typically treat all unmatched query-document pairs as equa...

📖 Read original article


29. Forecasting Revenue with its Customer-Base Drivers: When and Why Coordination Helps ​

Author: Kyeongbin Kim, Daniel McCarthy, Dokyun Lee
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02911v1 Announce Type: new Abstract: Revenue forecasts guide acquisition budgets, demand planning, and customer-based valuations, yet an aggregate forecast does not show whether change reflects acquisition, repeat purchasing, spending per order, or offsetting movements. Using weekly trans...

📖 Read original article


30. When Should Graph Attention Be Sparse? Learning a Per-Edge Tsallis Index ​

Author: Kleyton da Costa, Bernardo Modenesi
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02938v1 Announce Type: new Abstract: Graph attention normalizes neighborhood scores with softmax, the maximum-entropy choice under Shannon statistics. But homophilic and heterophilic graphs want different attention shapes, and one fixed normalization cannot serve both. We propose \textbf{...

📖 Read original article


31. Federated generative event models for tokenized electronic health records ​

Author: Michael C. Burkhart, Luke Solo, Inhyeok Lee, S'Khaja Charles, Zewei "Whiskey" Liao, Kaveri Chhikara, Dema Therese, Wan-Ting Liao, Catherine A. Gao, William F. Parker, Brett K. Beaulieu-Jones
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2608.02939v1 Announce Type: new Abstract: Electronic health record foundation models are limited by institutionally siloed data and substantial performance degradation under cross-site transfer. We evaluated federated training of tokenized generative event models (GEMs) across 122,251 intensiv...

📖 Read original article


32. Sedentary Behavior Classification for Wearable Sensors with a CNN-BiLSTM Model ​

Author: Yuliang Chen, Weiwei Shi, Jingjing Zou, Rong Zablocki, Animesh Kumar, Jordan A. Carlson, Sheri J. Hartman, Mikael Anne Greenwood-Hickman, Paul R. Hibbing, Marta Jankowska, Jay Yang, Arun Kumar, Loki Natarajan
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02946v1 Announce Type: new Abstract: Accurate detection of sedentary behavior is important for studying health risks related to prolonged sitting, but posture-based classification remains challenging with wearable sensors, especially at the wrist. We study whether a deep learning model tr...

📖 Read original article


33. ATFlash: Per-RoPE-Wavelength Attention Windows for Compute/Memory-Efficient LLM Inference ​

Author: Shun-ichiro Hayashi, Daichi Mukunoki, Tetsuya Hoshino, Takahiro Katagiri
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.02947v1 Announce Type: new Abstract: The attention score with rotary position embeddings (RoPE) decomposes exactly into a sum over its 2D-rotation frequency pairs, and each pair's wavelength limits how far it can discriminate position. Aligned with this structure, we propose the per-RoPE-...

📖 Read original article


34. Rubrics as Privileged Information for Open-Ended Generation ​

Author: Deepika Bablani, Ajay Gupta, Wanming Chen
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02948v1 Announce Type: new Abstract: On-policy self-distillation (OPSD), where a single model acts as both student and teacher with different contexts, has shown promise in verifiable domains like math, where hard privileged information (PI) in the form of ground-truth answers structurall...

📖 Read original article


35. Schedule-Informed Temporal Fusion Forecasting of Hourly Airport Security-Checkpoint Throughput ​

Author: Yinxiao Zhang, Sen Wang, Yi Gao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.ET

arXiv:2608.02950v1 Announce Type: new Abstract: Checkpoint staffing requires accurate forecasts of when screening demand will occur, yet flight schedules record departure times rather than passenger arrival times at security checkpoints. This study develops a framework that converts known flight sch...

📖 Read original article


36. SP3O: Reinforcement Learning from Segment Preferences without Reward Modeling ​

Author: Evan Assmus, Qining Zhang, Lei Ying
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.02951v1 Announce Type: new Abstract: Preference-based reinforcement learning (PbRL) for general stochastic MDPs often requires training a reward model. Existing reward-model-free methods are either restricted to bandits or deterministic MDPs, such as DPO or P3O, or use zeroth-order, gradi...

📖 Read original article


37. Inverted Detection and Control in Steering Vectors ​

Author: Max Torop, Aria Masoomi, Jennifer Dy
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02957v1 Announce Type: new Abstract: Steering vectors (SVs) are widely used to influence the expression of concepts (e.g., truthfulness) in large language model outputs. A key assumption underpinning SVs is that they are linearly discriminative with respect to the concept: representations...

📖 Read original article


38. Scaling an Autoregressive Transformer for Single-Cell Generation ​

Author: Aleksandr Sharipov, Yusif Mukhtarov, Igor Molybog
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.GN

arXiv:2608.02961v1 Announce Type: new Abstract: We study a self-supervised generation task for single-cell gene expression vectors: given a set of vectors from a cell type, we aim to generate additional gene expression vectors of that cell type. For this task we characterize both the biological fide...

📖 Read original article


39. A Physics-Informed Hybrid Neural Operator for Transient Magnetization Prediction in Power Magnetics ​

Author: Yachao Zhu, Qiujie Huang, Sinan Li, Yang Li, Gang Lei, Jianguo Zhu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.SY, eess.SY

arXiv:2608.02965v1 Announce Type: new Abstract: Magnetic components in high-frequency, high-power-density converters are increasingly driven by non-sinusoidal flux-density waveforms with fast transitions, minor-loop operation, dc bias, and temperature variation. Under these conditions, steady-state ...

📖 Read original article


40. Temporal Leakage in LLM Backtesting: Measurement, Validation, and Adjusted Scores ​

Author: Zeyu Zhang, Bradly C. Stadie
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, stat.ML

arXiv:2608.02985v1 Announce Type: new Abstract: The standard check for contamination in LLM backtests is simple: compare scores before and after the training cutoff. We show this check is uninformative. Four flagship models fail it on questions they cannot have memorized: every scored question resol...

📖 Read original article


41. AcceptMoE: Commitment-Weighted Self-Sizing Verifier Expert Sets for Efficient MoE Speculative Decoding ​

Author: Shuang Liang (Mark), Hao (Mark), Chen, Zhiwen Mo, Qianzhou Wang, Guoyu Li, Lingxiao Ma, Wayne Luk
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.DC

arXiv:2608.02989v1 Announce Type: new Abstract: Speculative decoding verifies a tree of draft tokens in one target-model forward pass. For a mixture-of-experts (MoE) target, however, parallel verification can activate the union of the experts selected by all tree nodes, even though only a small subs...

📖 Read original article


42. Joint Affine Spectral Shaping: Coupling Weight and Bias Updates Beyond Weight-Only Muon ​

Author: Gongyue Zhang, Honghai Liu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02991v1 Announce Type: new Abstract: Matrix spectral optimizers reshape weight-update spectra but usually delegate vector-valued biases to a separate optimizer. We study whether this separation is neutral. We formulate each affine layer as a joint momentum matrix $A=[M_W,\alpha m_b]$ and ...

📖 Read original article


43. A Graph Signal Processing Perspective on Numerical Sequence Representations in LLM In-Context Learning ​

Author: Jiajun Bao, Zihao Qi, Toni J. B. Liu, Gurbir Arora, Rapha"el Sarfati, Nicolas Boull'e, Christopher J. Earls
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP

arXiv:2608.03015v1 Announce Type: new Abstract: Pretrained large language models (LLMs) have demonstrated in-context learning (ICL) capabilities for numerical inference over sequences serialized as text. Prior work has identified and characterized this form of numerical inference primarily through o...

📖 Read original article


44. Paired Recipient-based Evaluation of Survival Prediction for Deceased Donor Kidney Transplants ​

Author: Misaki Matsuura, Mohammadreza Nemati, Dulat Bekbolsynov, Stanislaw Stepkowski, Kevin S. Xu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CY, stat.AP

arXiv:2608.03017v1 Announce Type: new Abstract: There has been significant interest in using machine learning algorithms to predict kidney transplant outcomes, such as the number of years until a graft inevitably fails. These prediction algorithms could possibly be used for pre-transplant donor-reci...

📖 Read original article


45. PLAN: Parallel Liquid-Inspired Approximation Network for Efficient Representation Learning in Flexible Job Shop Scheduling ​

Author: Dhivya Dharshini Kannan, Wei Zhang, Jieyi Bi, Yingpeng Du, Tianjun Wei, Jie Zhang, Zuming Liu, Anupam Trivedi
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03041v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) approaches for flexible job shop scheduling (FJSP) heavily rely on attention-centric architectures to achieve state-of-the-art performance. However, these models suffer from excessive parameter counts and prohibitive i...

📖 Read original article


46. Exploiting Separability in Multi-Scale Grey-Box Bayesian Optimization ​

Author: Joshua E. Hammond, Tyler A. Soderstrom, Brian A. Korgel, Michael Baldea
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03045v1 Announce Type: new Abstract: We consider grey-box optimization problems where the decision variables naturally partition into black-box variables (as arguments to an expensive black-box function) and white-box variables, governed by a set of explicit, closed-form equations that al...

📖 Read original article


47. Revisiting TD Target Aggregation under Uncertainty in Q-Learning ​

Author: Lipeng Zu, Xiaonan Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03069v1 Announce Type: new Abstract: Deep Q-Networks (DQNs) learn value functions through bootstrapped temporal-difference updates, where future returns are approximated using a greedy maximization over next-state action values. While effective, this aggregation rule is inherently sensiti...

📖 Read original article


48. SynEnergy: Anomaly Semantic-Guided Diffusion for Synthetic Energy Data Generation ​

Author: Lin Jiang, Dahai Yu, Ravikumar Gelli, Guang Wang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03087v1 Announce Type: new Abstract: Fine-grained energy consumption data are essential for applications such as demand forecasting, demand response planning, and grid reliability assessment. However, access to such data is often restricted by privacy concerns and data-sharing constraints...

📖 Read original article


49. SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation ​

Author: Wen Wang, Jiahua Bao, Tu Yongsiqi, Yihao Liu, Haotian Zhou, Haoxuan Ma, Mengyu Zhou, Wenkui Fan, Junwei He, Xiaoxi Jiang, Guanjun Jiang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.03092v1 Announce Type: new Abstract: We aim to improve model performance in multi-reward reinforcement learning training process. Existing Group reward-Decoupled Normalization Policy Optimization (GDPO) has mitigated the issue of reward signals masking one another during direct scalarizat...

📖 Read original article


50. Convex-Hull-Neighborhood Smooth Dual Generalization: Controlling Local Correction Propagation in Offline RL ​

Author: Yi Yang, Zhennan Chen, Mingfeng Lv, Hanlei Li, Zhengsen Ruan, Lvqing Yang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.03108v1 Announce Type: new Abstract: Offline reinforcement learning (offline RL) can benefit from nearby out-of-distribution (OOD) actions, but estimation errors at these actions may be amplified by bootstrapping. Existing regularization and local-generalization methods control either the...

📖 Read original article


51. Double Descent in Gradient Boosting Decision Trees via Split-Candidate Scaling ​

Author: Ryuichi Kanoh
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03111v1 Announce Type: new Abstract: Double descent is commonly studied by scaling an explicit capacity parameter, such as neural-network width. For gradient boosting decision trees (GBDTs), however, an analogous single-axis capacity parameter has not been established. We propose the numb...

📖 Read original article


52. Simulation-free and finite-time diffusion model ​

Author: Kentaro Kaba, Masayuki Ohzeki, Yuki Sughiyama
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03117v1 Announce Type: new Abstract: The performance of generative diffusion models is determined by the choice of the reference diffusion process connecting the empirical and prior distributions. Conventional approaches typically trade off simulation-free training against finite-time gen...

📖 Read original article


53. Trajectory-Guided Forget-Recover Network for Continual LLM Unlearning ​

Author: Zezheng Wu, Xinghe Cheng, Qinggang Zhang, Haoran Luo, Jiapu Wang, Qing Yang, Jingwei Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03123v1 Announce Type: new Abstract: Machine unlearning aims to eliminate the influence of sensitive data on a model. In the real world, unlearning requests arrive continually, which gives rise to two challenges. First, an unlearning intervention may redistribute target-related computatio...

📖 Read original article


54. Lightweight Chunk Selection for Mobile Retrieval-Augmented Generation ​

Author: Sicong Chang, Yidan Shen, Wen Yu, Jiefu Chen, Xin Fu, Renjie Hu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03148v1 Announce Type: new Abstract: RAG improves the factual grounding of LLM by incorporating external knowledge, but deploying RAG on mobile and edge devices remains challenging because retrieved context increases computation and memory. A direct way to reduce this cost is to retain on...

📖 Read original article


55. On the Implicit Flatness Bias of Sharpness-Aware Minimization: A Linear Stability Analysis with Quantitative Hyperparameter Bounds ​

Author: Jiaxin Deng, Junbiao Pang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03197v1 Announce Type: new Abstract: Sharpness-Aware Minimization (SAM) improves generalization by seeking parameters whose loss is robust to local adversarial perturbations, but the quantitative mechanism underlying its implicit bias toward flat minima remains unclear. In particular, the...

📖 Read original article


56. Agentic Reinforcement Learning with Self-Distilled Reward Shaping ​

Author: Ranxu Zhang, Guinan Chen, Chenshaodong, Jinghao Lin, Xiaozhou Xu, Sunzhe, Yanyong Zhang, Chao Wang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.03223v1 Announce Type: new Abstract: Agentic reinforcement learning enables LLM agents to learn through interaction, but sparse trajectory-level rewards reveal success without identifying which intermediate decisions deserve credit. Training-only privileged skills can provide denser super...

📖 Read original article


57. SAKI: Score-Aware Low-Rank Key Indexing for Long-Context KV Retrieval ​

Author: Lin Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, stat.AP

arXiv:2608.03228v1 Announce Type: new Abstract: Existing low rank KV cache methods preserve either model weights or key variance, neither of which directly reflects the attention scores used during inference. We derive the expected attention score distortion caused by rank r key compression and show...

📖 Read original article


58. FinVerse: Financial Time-Series Benchmark ​

Author: Jaehoon Lee, Jun Seo, Seunghan Lee, Tae Yoon Lim, Dongwan Kang, Hwanil Choi, Minjae Kim, Sungdong Yoo, Junhyeok Kang, Sangjun Han, Soonyoung Lee, Wonbin Ahn
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03259v1 Announce Type: new Abstract: As time-series foundation models have emerged, the need for benchmarks that can evaluate their forecasting ability in meaningful ways has become increasingly important. Existing time-series forecasting benchmarks provide useful standardized comparisons...

📖 Read original article


59. ED-DiT: Physics-Guided Diffusion Pretraining for Transferable Molecular Representations from Electron Density ​

Author: Liang Shuang, Haocheng Wang, Jiayi Song, Shuquan Ye, Ben Fei
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03260v1 Announce Type: new Abstract: Pretraining has shown strong potential for learning transferable representations, yet it remains underexplored for electron-density-based molecular learning. Electron density provides a continuous three-dimensional description of molecular electronic s...

📖 Read original article


60. The Ignition Is Real, and It Lives at the Readout: Latent composition, difficulty-clocked ignition, and the interface-constituted commit in a recurrent-depth reasoner ​

Author: Simon Lam-Muir
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03263v1 Announce Type: new Abstract: We test whether the "compositional ignition" reported in latent-reasoning models is real computation, an instrument artifact, or inherited from verbal training data. We grow an independent realization of a published 30M-parameter recurrent-depth reason...

📖 Read original article


61. Noise-Aware Shrinkage for Differentially Private Zeroth-Order Fine-Tuning of Large Language Models ​

Author: Lele Zheng, Weifeng Kong, Xinyi Zhang, Ke Cheng, Tao Zhang, Yulong Shen
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.03277v1 Announce Type: new Abstract: Differentially private zeroth-order optimization (DP-ZO) enables memory-efficient private fine-tuning of large language models using only forward evaluations. Existing aggregation-based DP-ZO methods reconstruct model updates at a fixed scale, ignoring...

📖 Read original article


62. The Tell-Tale Trace: Detecting Reasoning Failures in LLMs Using Chain-of-Thought Dynamics ​

Author: Shashwat Sourav, Aishwarya Balwani
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.03291v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning improves large language model (LLM) performance while also providing an observable interface to the model's reasoning process. Existing approaches that leverage verbalized CoTs to monitor reasoning correctness, however,...

📖 Read original article


63. Provably Learning Multi-Head Attention with Queries ​

Author: Sunyeop Kim, Insung Kim, Jian Guo
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.03294v1 Announce Type: new Abstract: We study the problem of learning multi-head softmax attention from black-box input-output access. The learner may query arbitrary real-valued token sequences and observe only the scalar output at the final token. Recent work gives an algorithm using $O...

📖 Read original article


64. Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging ​

Author: Siming Fu, Zheming Fu, Ruizhe He, Hualiang Wang, Jie Huang, Xiaoxiao Ma, Mingchen Zhong, Weihu Huang, Xiaoxuan He, Haojun Xu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.03316v1 Announce Type: new Abstract: On-policy distillation, in which a teacher corrects samples that the student itself generates, presupposes that the two models speak the same language: identical VAE latents, matching architectures, and a common timestep grid. We ask what happens when ...

📖 Read original article


65. AS-FedBridge: Pseudo-Spike Bridge Distillation for Heterogeneous ANN-SNN Federated Learning ​

Author: Shengyang Li, Yiting Dong, Liuyang Song, Ximing Wang, Luyuan Xie, Cong Li, Qingni Shen, Zhaofei Yu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2608.03324v1 Announce Type: new Abstract: Federated learning enables collaborative model training across distributed edge devices while strictly preserving data privacy. To facilitate practical deployment on resource-constrained edge devices, Spiking Neural Networks (SNNs) have emerged as a pr...

📖 Read original article


66. Tight Worst-Case Bounds for the Smallest Eigenvalue of ReLU NTK Gram Matrices ​

Author: Zhao Song
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03368v1 Announce Type: new Abstract: For $n$ unit vectors $x_1,\ldots,x_n \in \mathbb{R}^d$, we study the continuous ReLU derivative Gram matrix $H$, whose entries are obtained by averaging pairwise gated inner products over a standard Gaussian direction. Writing $ \Delta_\pm := \min_{i ...

📖 Read original article


67. Benign interpolation and Occam's razor ​

Author: Tom F. Sterkenburg, Daniel A. Herrmann, Jan-Willem Romeijn
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03386v1 Announce Type: new Abstract: Contemporary deep learning methods generalize well even when they fit their training data perfectly, a phenomenon known as benign interpolation. This phenomenon cannot be accounted for by classical statistical learning theory and has prompted a range o...

📖 Read original article


68. TimeRLM: Recursive Language Models Enable Precise Anomaly Localization in Long-Context Time-Series ​

Author: Nicolas Zumarraga, Lorenzo Steno, Ning Wang, Max Rosenblattl, Thomas Kaar, Maxwell A. Xu, Kevin O'Sullivan, Markus Kreft, Elgar Fleisch, Paul Schmiedmayer, Patrick Langer, Robert Jakob
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03391v1 Announce Type: new Abstract: Precise anomaly localization over long-context time series is a crucial task in monitoring applications across clinical care, industrial operations, financial services, and logistics, where brief evidence may hide inside long spans of high-frequency da...

📖 Read original article


69. Shorter Reasoning, Earlier Answers? An Evaluation of Reasoning Interfaces ​

Author: Francesca Carlon, Vincent Ginis, Andres Algaba
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03401v1 Announce Type: new Abstract: Large language models often reason at length before answering, increasing cost and latency. Prompts and trained settings can shorten this reasoning, but a shorter trace may only show that the model stopped sooner. Here, we evaluate paired runs of the s...

📖 Read original article


70. Stop Replacing Noise with Noise: Two-Source Reliability Assessment for Label Correction and Sample Reweighting in Label-Noise Learning ​

Author: Wenxiao Fan, Kan Li
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.03432v1 Announce Type: new Abstract: Refurbishment-based noisy-label learning mixes an observed label with a model-derived pseudo target, typically using one sample-wise cleanliness score to control both branches. This creates a hidden coupling: reducing trust in the observed label automa...

📖 Read original article


71. Approximate Speculative Decoding ​

Author: Yuannuo Feng, Zegang Peng, Yuxin Xie, Yubing Ye, Yizhe Chen, Wenshuai Yao, Wenyong Zhou, Wang Kang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03447v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by verifying a draft block with a target model in parallel. Under standard greedy verification, decoding stops at the first draft token that differs from the target argmax, discarding the remai...

📖 Read original article


72. Beyond the Gegenbauer Paradigm: q-Orthogonal Kernels for Machine Learning ​

Author: 'Alvaro S'anchez-Paniagua R'ios, Juan P. Llerena, Alberto Lastra, Nuria Torrado, Edmundo J. Huertas
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.03482v1 Announce Type: new Abstract: The performance of Support Vector Machines (SVMs) critically depends on the kernel function choice, which enables implicit mapping of data into high-dimensional feature spaces. While classical kernels like Radial Basis Function (RBF) remain popular, or...

📖 Read original article


73. FedCARE: A Multi-Objective Personalised Federated Learning Framework for Smart Healthcare ​

Author: Rojalini Tripathy, Padmalochan Bera, Shreya Ghosh, Rajkumar Buyya
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2608.03498v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative model training across distributed healthcare institutions without centralising sensitive patient data. However, real-world healthcare federations are often characterised not only by non-IID data, but also b...

📖 Read original article


74. Robust General Utility for Reinforcement Learning ​

Author: Zixuan Liu, Fangzheng Wu, Brian Summa, Zizhan Zheng
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03562v1 Announce Type: new Abstract: Reinforcement learning (RL) with general utility extends classic RL by optimizing an arbitrary utility functional of the policy-induced occupancy measure, thereby enabling a broader range of applications. However, previous work on general utility RL ty...

📖 Read original article


75. Pin Once, Swap Light: Subspace-Aligned Centroid-Residual Training for Efficient Ultra-LoRA Serving ​

Author: Xiang Li, Pengcheng Wang, Huazheng Wang, Saurabh Bagchi
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03579v1 Announce Type: new Abstract: Modern multi-tenant Low-Rank Adapters (LoRAs) serving systems concurrently host tens to hundreds of LoRA adapters. Though powerful, this introduces a critical system dilemma between serving efficiency and task performance: higher-rank adapters generall...

📖 Read original article


76. Design-Time Optimization of Deep Neural Networks for Intermittent Learning on Microcontrollers ​

Author: Jakob Schubert, Maximilian Kasper, Maximilian Linke, Benedict Herzog, Mark Deutel, Axel Plinge, Dominik Seuss, Christopher Mutschler
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03589v1 Announce Type: new Abstract: We present a method for designing deep neural networks (DNNs) for intermittent, energy-autonomous, on-device learning on microcontroller units (MCUs). In mobile applications where the energy can run out, e.g., when solar-powered, executing artificial i...

📖 Read original article


77. A Theory of Conditional Collapse under Low-Rank Weight-Space Ablations: I. The Single-Block Theory and Synthetic Validation ​

Author: Abdallah Khemais
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03620v1 Announce Type: new Abstract: Activation patching and weight-space ablation both claim a component is causally responsible for a behavior, yet they act on different objects: one forward pass versus the parameters behind every forward pass. We ask when they agree. We study an ideali...

📖 Read original article


78. ConformalShift: Targeted Event Reordering Against Adaptive ECG Monitoring ​

Author: Arash Vashagh, Yasmin Vashagh
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03628v1 Announce Type: new Abstract: Adaptive conformal prediction can recover clinically important heartbeat classes missed by a point classifier, but delayed feedback makes its decisions sensitive to event order. We introduce ConformalShift, a bounded event-reordering attack that suppre...

📖 Read original article


79. POEM: Phase-Aware $\mathrm{SO}(2)$ Feature Rotation for Time Series Forecasting Under Periodicity Drift ​

Author: Jiawen Zhu, Shuhan Liu, Shengxuan Li, Qiming Shi, Di Weng
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03630v1 Announce Type: new Abstract: Deep learning has advanced time series forecasting, but periodicity drift, in which cycle timing and phase vary over time, remains a challenging problem. Existing methods predominantly model these sequences on fixed time grids, suffering from a limited...

📖 Read original article


80. CausalOPD: First-Wrong-Step Supervision for Distilling Causal Chain Reasoning ​

Author: Jian Zhang, Bingyi Wang, Yizhi Liu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03673v1 Announce Type: new Abstract: Many critical reasoning tasks, including clinical diagnosis, legal judgment, and industrial fault diagnosis, require step-dependent causal chains in which early errors propagate and correct conclusions can mask invalid reasoning. Although large languag...

📖 Read original article


81. DiagLoop: A Counterfactual Data Flywheel with Stage-Localized Reinforcement for Diagnostic LLMs ​

Author: Jian Zhang, Bingyi Wang, Yizhi Liu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03674v1 Announce Type: new Abstract: Causal diagnostic models must explain how conclusions follow from evidence because diagnoses guide repairs and treatments. Yet serious cases are scarce, records rarely contain reasoning paths, and data transfer poorly across configurations, complicatin...

📖 Read original article


82. LAEF: A Lead-Agnostic ECG Foundation Model Towards Point-of-Care Diagnostics ​

Author: Edoardo Coppola, Stefano Fiorini, Pietro Li`o, Mattia Savardi, Alberto Signoroni
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03690v1 Announce Type: new Abstract: Point-of-care cardiac devices such as smartwatches and handheld ECG recorders typically capture 1--2 leads, yet existing ECG foundation models are architecturally constrained to fixed 12-lead inputs, degrading or failing under these reduced configurati...

📖 Read original article


83. Learning and Clustering on Temporal Graphs: Principles, Primitives, and Pooling ​

Author: Nelson Aloysio Reis de Almeida Passos, Emanuele Carlini, Salvatore Trani
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.SI

arXiv:2608.03696v1 Announce Type: new Abstract: This work focuses on the problem of learning on temporal graphs, with particular emphasis on the task of clustering: obtaining coarse-grained representations by aggregating information from nodes, edges, and temporal dynamics - a task related to poolin...

📖 Read original article


84. To Describe or Construct Statistical Learning Models Using the Category-theoretical Language ​

Author: Congwei Song
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03706v1 Announce Type: new Abstract: Statistical learning is a fascinating field that has long been the mainstream of machine learning/artificial intelligence. A large number of results have been produced which can be widely applied to real-world problems. It also leads to many research t...

📖 Read original article


85. Amortized Interventional Forecasting for Multivariate CIR Processes ​

Author: Andreas Sauter, Sumit Sourabh, Drona Kandhai, Erman Acar
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CE

arXiv:2608.03715v1 Announce Type: new Abstract: Mean-reverting dynamics are pervasive in finance, and the Cox--Ingersoll--Ross (CIR) process is a standard model for the time series they produce, from short rates to credit default swap (CDS) spreads. Yet CIR models capture only \emph{correlated} co-m...

📖 Read original article


86. UNVaMP: Neural Knowledge Tracing with Variational Regularization of Latent Knowledge Dynamics ​

Author: Carson J. Cook, Ahmed J. Zerouali, Anthony Schmidt, Reginald Ziedzor, Paul Lin, Luke G. Eglington
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2608.03811v1 Announce Type: new Abstract: We introduce the Unified Neural Variational Measurement of Proficiency (UNVaMP) architecture, a knowledge tracing method that integrates observed student-item interactions with internal memory to produce evolving latent representations of student knowl...

📖 Read original article


87. Resume Means Resume: A Machine-Checked Conformance Contract for Checkpoint, Interrupt, and Resume Semantics in Workflow Persistence Layers ​

Author: Sajjad Khan
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.LO, cs.SE

arXiv:2608.03836v1 Announce Type: new Abstract: A framework that persists execution state so a run can be interrupted, survive a crash, and continue must decide what a resume means for effects that already fired. Five widely deployed agent workflow frameworks answer differently, none exposes a machi...

📖 Read original article


88. FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs ​

Author: Amin Farajzadeh, Melike Erol-Kantarci
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.MA, cs.NI

arXiv:2608.03852v1 Announce Type: new Abstract: This paper proposes FedCritic-MIMO, a communication-efficient serverless federated multi-agent reinforcement learning framework for AI-native resource control across independently deployable cell-level controllers in open and disaggregated 6G RANs. Con...

📖 Read original article


89. Quantization Effects on Biomedical LLM Reliability ​

Author: Anton Rasmussen, Hong Qin
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03854v2 Announce Type: new Abstract: When decoder language models are used as classifiers, predicted class probabilities depend on implementation choices, including the prompt template, verbalizer (label-to-token mapping), and scoring rule, that are rarely treated as experimental variable...

📖 Read original article


90. Bi-semantic Chemical Embedder for Joint Representation Learning of SMILES and Natural Language ​

Author: David Ming Segura, Jeremy Goumaz, Joshua W. Sin, Bojana Rankovi'c, Philippe Schwaller
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03855v1 Announce Type: new Abstract: Transformer models have revolutionized natural language processing (NLP), and text-based molecular representations like SMILES have successfully extended these architectures to chemistry. However, domain-adaptive pre-training often causes models to ove...

📖 Read original article


91. CRS-Triage: Confidence- and Reliability-Aware Selective Triage under Incomplete Clinical Evidence ​

Author: Guan Qiang, Yushen Chen, Tianlong Liu, David Rotenberg, Ethan H. Kim, Fang Fang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03862v1 Announce Type: new Abstract: Emergency triage requires reliable decisions within a short time period. However, the available electronic health record (EHR) data, including structured data and clinical text, are often incomplete, unreliable, and inconsistent. This makes machine lea...

📖 Read original article


92. GENESIS: Towards Explainable Causal Discovery ​

Author: Abhinav Thorat, Ravi Kumar Kolla, Vishak K Bhat, Harsh Vardhan Singh Chauhan, Niranjan Pedanekar
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03868v1 Announce Type: new Abstract: Causal Discovery (CD) from observational data faces two fundamental challenges. First, purely statistical methods often lack the power to resolve structural ambiguities in low-sample regimes. Second, although LLM-assisted hybrid approaches improve stru...

📖 Read original article


93. Enhancing VLM Reward Models Through Structure-Aware Fine-Tuning ​

Author: Pyrros Koussios, Chenhao Li, Xin Chen, Andreas Krause
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03875v1 Announce Type: new Abstract: Designing effective reward functions remains a major bottleneck in Reinforcement Learning (RL). Recent work uses large foundation Vision-Language Models (VLMs) as reward models, computing text-observation similarity to bypass manual reward engineering....

📖 Read original article


94. Operationally Feasible Synthetic Power-Grid Scenarios via Learning the AC-Operable Joint Distribution ​

Author: Chenhan Xiao, Xinyu He, Haoran Li, Hanghang Tong, Yang Weng
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2608.03878v1 Announce Type: new Abstract: Synthetic power-grid scenarios are essential for planning, resilience assessment, contingency analysis, and data-driven power-system applications. Recent synthetic grid generation methods have improved structural realism and operational feasibility by ...

📖 Read original article


95. Omega-S: A Functional Resilience Index for LLM Fine-Tuning ​

Author: Alberto Acedo
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, q-bio.MN

arXiv:2608.03887v1 Announce Type: new Abstract: Fine-tuning a large language model on new data degrades what it previously learned. We present Omega-S, a drop-in penalty computed from the weight matrix alone: it needs no previous-task data, no Fisher matrix and no stored copy of the old weights. It ...

📖 Read original article


96. Cross-Model KV Cache Transfer in LLM Families: A Closed-Form Linear Mapping for Prefill Reuse ​

Author: Taekyung Heo, Rasoul Shafipour, Ritchie Zhao, Maximilian Golub, Mohammad Mahdi Kamani, Ritika Borkar, Makesh Tarun Chandran, Pantea Zardoshti, Bita Darvish Rouhani
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03893v1 Announce Type: new Abstract: Production deployments often swap between different-sized models in a family for cost-quality cascading, mid-conversation switching, and routing, and each swap forces the receiver to repay the prefill from scratch. We propose cross-model KV cache trans...

📖 Read original article


97. Sparse Weight Decomposition for Efficient Circuit Extraction ​

Author: Chuanhao Yan, Xuhan Huang, Yawen Duan, Zhenfei Yin, Hang Zhao, Bryan Dai, Jie Fu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.03913v1 Announce Type: new Abstract: Dense pretrained transformers do not naturally expose interpretable units for circuit extraction. Existing approaches obtain such units by learning auxiliary sparse representations or training sparse models, incurring substantial additional computation...

📖 Read original article


98. Trajectory inference via Acceleration Matching ​

Author: Bartolo Dazzini, Giovanni Conforti, Alain Durmus, Aram-Alexandre Pooladian
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2608.03916v1 Announce Type: new Abstract: Trajectory inference is a fundamental problem in many scientific domains: given a collection of unpaired snapshots of observations at discrete time points, the goal is to generate smooth trajectories that best resemble and interpolate the data. Existin...

📖 Read original article


99. PRISM: Powerful Time Series to Image (TS2I) Representations for Multivariate Anomaly Detection ​

Author: Mateusz Smendowski, Kamil Faber, Piotr Nawrocki, Nathalie Japkowicz, Roberto Corizzo
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.03926v1 Announce Type: new Abstract: Time series anomaly detection (TSAD) underpins applications in predictive maintenance, finance, and cloud computing, however performance remains sensitive to representation choices, especially in multivariate settings. While transforming time series in...

📖 Read original article


100. A Physics-Flavored Transformer Network for Parametrizing Contraction Dynamics of Engineered Skeletal Muscle Tissues ​

Author: Mattias Luber, Timo Betz
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03927v1 Announce Type: new Abstract: Engineered Skeletal Muscle Tissues (ESMs) have become a key structure for biomedical disease modeling and pharmacological screening, yet their functional characterization often relies on simplistic metrics like peak force, discarding critical kinetic i...

📖 Read original article


101. Latent Reward Registers for Diffusion Preference Alignment ​

Author: Yuanshen Guan, Zipeng Feng, Chengru Song, Zhiwei Xiong, Peiqin Sun
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.03929v2 Announce Type: new Abstract: Aligning diffusion models with human preferences usually relies on a sparse terminal reward evaluated on the final generated samples, presenting a severe temporal credit-assignment challenge across the multi-step denoising process. We propose Latent Re...

📖 Read original article


102. Muon Meets Mamba: Spectral Optimization for State Space Models ​

Author: Arslan Battalov, Karim Kramin, Alexander Markotenko, Sofia Sinitsina
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03941v1 Announce Type: new Abstract: Muon is a recent optimizer that orthogonalizes the update to each weight matrix with a Newton-Schulz iteration, which performs steepest descent under the spectral norm. Almost all the evidence for it comes from Transformer models, and its behavior on s...

📖 Read original article


103. Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation ​

Author: Seyed Kahaki, Shijie Li, Weijie Chen, Nicholas Petrick
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03990v1 Announce Type: new Abstract: Synthetic histopathology image generation has emerged as an approach that may address data scarcity in computational pathology, yet current evaluation methodologies may not fully assess synthetic data quality for medical applications. This work investi...

📖 Read original article


104. Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility ​

Author: Mohsen Hariri, Weicong Chen, Nahal Shahini, Vikash Singh, Kai Ye, Amirhossein Samandar, Debargha Ganguly, Sreehari Sankar, Yanyan Zhang, Shouren Wang, Jerry Peng, Biyao Zhang, Michael Hinczewski, Vipin Chaudhary
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04001v1 Announce Type: new Abstract: Large language models can solve substantially harder reasoning problems with more inference-time compute. The term "test-time scaling," however, now covers diverse inference algorithms that extend deliberation along a single trajectory, sample complete...

📖 Read original article


105. AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks? ​

Author: Dong Yan, Jian Liang, Dapeng Hu, Ran He, Nicholas Jing Yuan, Qi Zhang, Tieniu Tan
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.00155v1 Announce Type: cross Abstract: Large language model (LLM) agents can self-evolve by continually improving from their own accumulated experience. However, existing studies predominantly adopt independent evaluation. Consequently, the behavior of self-evolving agents in realistic st...

📖 Read original article


106. TabletCraft: Bridging a 4,000-Year Cultural Gap with Bidirectional Akkadian NMT and Cuneiform Rendering ​

Author: Zhaohui Wang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.02609v1 Announce Type: cross Abstract: Half a million cuneiform clay tablets survive in museums worldwide, yet modern users can neither read nor write in the world's oldest writing system, leaving a 4,000-year cultural barrier that existing NLP tools have only partially addressed. Prior w...

📖 Read original article


107. KernelBrain: Coarse-to-Fine, Budget-Aware Search for Agentic GPU Kernel Optimization ​

Author: Shuai Che, Gang Peng
Published: 8/5/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG, cs.SE

arXiv:2608.02611v1 Announce Type: cross Abstract: Automating GPU kernel optimization remains difficult in practice: generated variants can violate correctness constraints, runtime measurements are noisy, and search often stalls early. We present a practical optimization agent that combines LLM-guide...

📖 Read original article


108. BBOWP-Bench: Evaluating LLMs on Black-Box Optimization Word Problems ​

Author: Yutaro Yamada, Kei Hiroshima, Nozomu Yoshinari, Kento Uchida, Shinichi Shirakawa
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.NE

arXiv:2608.02612v1 Announce Type: cross Abstract: Formulating an optimization problem strongly affects the quality of the final solution, yet good formulations usually require substantial expertise. Recent studies have therefore examined how to automatically derive optimization problems from natural...

📖 Read original article


109. MemArena: An Ego-Centric Benchmark for On-Device Agentic Personal Memory Assistants at Scale ​

Author: Jiadong Zhang, Xiaosong Ma
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.MA

arXiv:2608.02613v1 Announce Type: cross Abstract: Edge-deployed personal memory assistants must handle private interpersonal conversations on-device with open-weight models. Yet, existing memory benchmarks often under-test the combination of activity-dense interaction, ego-centric perspective, and c...

📖 Read original article


110. Preferred, Not Safer: Pairwise Preference Is a Poor Proxy for Clinical Safety ​

Author: Fay Elhassan, David Sasu, Alexandra Kulinkina, Lars Henning Klein, Mary-Anne Hartley
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.02617v1 Announce Type: cross Abstract: We evaluate whether clinician pairwise preferences provide a reliable signal of clinical safety in large language model (LLM) evaluation using expert feedback from MOOVE (Massive Open Online Validation and Evaluation), a clinician-led platform collec...

📖 Read original article


111. Neural network realization of binary refinement iterates via a two-chart atlas selector ​

Author: Tsogtgerel Gantumur
Published: 8/5/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, stat.ML

arXiv:2608.02624v1 Announce Type: cross Abstract: Refinement operators generate many functions used in wavelet constructions, subdivision schemes, and geometric modeling. Their finite iterates can develop rapidly increasing numbers of linear pieces, making them a natural test case for the expressive...

📖 Read original article


112. Micro-Segmentation Anomaly Detection in Zero-Trust Software-Defined Network Fabrics ​

Author: Ashly Joseph
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.CV, cs.LG, cs.NI

arXiv:2608.02627v1 Announce Type: cross Abstract: Zero Trust Architecture (ZTA) principles need rigorous network segmentation and ongoing verification to reduce implicit trust and lateral threat propagation. This paper investigates anomaly detection in software-defined networking (SDN) systems by mi...

📖 Read original article


113. On the Performance of Malware Detection Classifiers Using Hardware Performance Counters ​

Author: Alireza Abolhasani Zeraatkar, Parnian Shabani Kamran, Inderpreet Kaur, Nagabindu Ramu, Tyler Sheaves, Hussain Al-Asaad
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.02671v1 Announce Type: cross Abstract: Malware detection using Hardware Performance Counters (HPC) has emerged as a promising solution to improve the security of computing systems as a complement to antivirus software. Hardware-based malware detectors (HMD) use Machine Learning (ML) class...

📖 Read original article


114. DenialRAG: Single-Document RAG Poisoning via Embedded Parametric Denial ​

Author: Abay Zhurekbay, Tao Liu, Fan Li
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.02678v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to corpus poisoning: an attacker who inserts a crafted document into the retrieval corpus can steer the underlying large language model (LLM) toward an attacker-chosen wrong answer. Prior si...

📖 Read original article


115. TraceCompiler: Skill-Guided Mining and Compilation of LLM Agent Traces into Mostly Deterministic Workflows ​

Author: Salma El Yadouni (EPFL), Guanyi Li (Binome Technologies)
Published: 8/5/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2608.02680v1 Announce Type: cross Abstract: Tool-using language-model agents repeatedly rediscover procedures they have already executed, producing traces that mix reusable structure with retries, exploration, accidental ordering, and repeated lookups. We present TraceCompiler, a skill-guided ...

📖 Read original article


116. Stuck on "A": Diagnosing and Repairing Interface Injury in Attention-to-KDA Linearization of a 0.6B Language Model ​

Author: Ronglong Bao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.02689v1 Announce Type: cross Abstract: We convert 21 of 28 full-attention layers of Qwen3-0.6B-Base into KDA (Kimi Delta Attention) linear-attention layers on a single consumer-grade GPU budget, and ask a simple question: what exactly does the conversion break? After surgery, hidden-state...

📖 Read original article


117. Crayotter: Learning Long-Horizon Video Editing Agents via Group-Relative Preference Backpropagation ​

Author: Lecheng Yan, Jianze Lin, Yichong Zhang, Ben Pan, Wenxi Li, Chenyang Lyu, Liting Zhou, Cathal Gurrin
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.02694v1 Announce Type: cross Abstract: Long-horizon video editing agents receive final-product feedback only after many interdependent decisions. Yet editing quality is subjective, admits multiple valid solutions, and is not meaningfully calibrated across heterogeneous requests, making a ...

📖 Read original article


118. Stylometric Defenses Against Author Impersonation in Software Repositories ​

Author: Leonid Ravich, Michael Fire
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.SE

arXiv:2608.02695v1 Announce Type: cross Abstract: Software supply-chain attacks increasingly exploit an identity gap where compromised maintainer accounts authorize malicious changes. This work evaluates patch-level authorship verification as a behavioral defense layer, showing that stylometric anal...

📖 Read original article


119. Explainable AI for the EU Right to Explanation: A Systematic Review of the Law-XAI Translation Gap ​

Author: Benjamin Fresz, Elena Dubovitskaya, Marco F. Huber
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG

arXiv:2608.02699v1 Announce Type: cross Abstract: When algorithms make or influence consequential decisions---about loan eligibility, hiring, or healthcare---EU law grants affected individuals a Right to Explanation. Yet whether (and how) Explainable AI (XAI) can satisfy this right in practice remai...

📖 Read original article


120. ARCHead: Activation-Metric Residual Correction for Large Language Model Output Heads ​

Author: \c{S}uayp Talha Kocabay, Talha R"uzgar Akku\c{s}, Kamer Ali Yuksel
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.02703v1 Announce Type: cross Abstract: Weight-only quantization substantially reduces the storage of large language model (LLM) transformer blocks, but practical backends often retain the final language-modeling head (LM-head) in BF16 or FP16. Quantizing this projection naively can strong...

📖 Read original article


121. DAIF: A Data-Driven Intermediate Fusion Framework for Multimodal Supervised Learning via Approximate Message Passing ​

Author: Sagnik Nandy, Samriddha Lahiry, Pragya Sur, Subhabrata Sen
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, math.ST, stat.ML, stat.TH

arXiv:2608.02769v1 Announce Type: cross Abstract: Multimodal supervised learning seeks to leverage multiple heterogeneous data sources to improve predictive performance. A central challenge is determining the fusion granularity across modalities: over-integration may amplify noise while under-integr...

📖 Read original article


122. A Hyperfinite Framework for Score-Based Generative Modeling ​

Author: Sunder Ram Krishnan
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, math.PR

arXiv:2608.02799v1 Announce Type: cross Abstract: Score-based diffusion models are typically formulated using continuous-time stochastic differential equations and measure-theoretic stochastic calculus. In this paper, we develop a hyperfinite formulation of score-based generative modeling within the...

📖 Read original article


123. Detecting high-frequency brain disorder signals using dynamic mode decomposition from EEG ​

Author: Jacob Kang, Jong-Hyeon Seo
Published: 8/5/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG, eess.SP

arXiv:2608.02804v1 Announce Type: cross Abstract: Recent studies have reported clearly identifiable dynamical changes in the high-frequency range of EEG signals recorded during specific stimuli, such as visual or auditory inputs, or in cases of brain disorders like epileptic seizures. In this study,...

📖 Read original article


124. Evading Chain-of-Thought Monitoring Through Model Poisoning ​

Author: Giorgio Severi, Shujaat Mirza, Blake Bullwinkel, Amanda Minnich
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.02820v1 Announce Type: cross Abstract: Chain-of-thought (CoT) monitoring is an increasingly important component of AI safety stacks but relies on the assumption that a model's reasoning trace is informative about its actions. This work studies the limits of CoT monitoring through the lens...

📖 Read original article


125. Improved Quantum Algorithms for Reinforcement Learning Under a Generative Model ​

Author: Joao F. Doriguello
Published: 8/5/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG, stat.ML

arXiv:2608.02826v1 Announce Type: cross Abstract: Reinforcement learning is a subfield of machine learning that studies how an agent interacts with an environment in order to extract as large a reward as possible. A standard approach to study such interaction is through Markov Decision Processes (MD...

📖 Read original article


126. Particle-based Generalised Stochastic Optimisation ​

Author: Jiechen Jackie Zhang, O. Deniz Akyildiz
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO

arXiv:2608.02844v1 Announce Type: cross Abstract: We develop a class of diffusion-based stochastic particle optimisation methods for loss functions with intractable gradients. Specifically, we consider problems in which the loss gradient is an integral with respect to a parameter-dependent distribut...

📖 Read original article


127. Field Aware Agent Skill Retrieval ​

Author: Paimon Goulart, Liang Wu, Kelly Wan, Evangelos E. Papalexakis, Liangjie Hong
Published: 8/5/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.02880v1 Announce Type: cross Abstract: As lifelong learning agents accumulate lifelong growing skill banks, retrieving the correct skill becomes an increasingly important bottleneck. Most current skill retrieval methods treat each skill as one flat document by concatenating fields such as...

📖 Read original article


128. ScoreField: Neural Inverse Scattering with Score-Based Generative Priors ​

Author: Wenhan Guo, Yuan Gao, Yu Sun
Published: 8/5/2026, 4:00:00 AM
Categories: eess.IV, cs.LG, physics.comp-ph

arXiv:2608.02937v1 Announce Type: cross Abstract: Designing an effective electromagnetic inverse-scattering solver requires faithful enforcement of nonlinear full-wave physics together with an expressive prior on the unknown permittivity contrast. We propose ScoreField, a neural inverse scattering f...

📖 Read original article


129. TQLite: Multi-LLM Jury Guided Distillation for Real-time MQM Translation Quality Evaluation ​

Author: Bhavin Jawade, Cameron R. Wolfe
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.02975v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated impressive performance in MQM-based translation quality (TQ) evaluation, and recent advances in large reasoning models (LRMs) promise even greater improvements. However, both LLMs and LRMs are computatio...

📖 Read original article


130. Stochastic Saddle Avoidance Beyond Unit Excitation and Smoothness: A Pathwise Lyapunov-Perron Framework ​

Author: Junwen Qiu, Bohao Ma, Andre Milzarek, Junyu Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DS, stat.ML

arXiv:2608.03001v1 Announce Type: cross Abstract: Unit excitation (UE) is a common assumption in stochastic saddle avoidance: the stochastic error must have a uniformly positive component along every direction, in expectation. This condition gives a direct way to rule out convergence to strict saddl...

📖 Read original article


131. LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs ​

Author: Forough Majidi, Mohammad Mehdi Morovati, Foutse Khomh, Heng Li
Published: 8/5/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2608.03036v1 Announce Type: cross Abstract: Large Language Models (LLMs) are integrated into software systems and AI services, making efficient LLM serving a concern for software engineering. Serving LLMs is challenging because inference requires computation, memory, GPU resources, and executi...

📖 Read original article


132. CVPO: Enhancing LLM Reinforcement Learning Reasoning via Value-Variance Adaptation and Dynamic Curriculum Learning ​

Author: Ziqi Jia, Yalu Ouyang, Bo Pang, Panpan Li, Hangfei Xu, Shengzhao Wen, Shiyong Li, Yanpeng Wang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.03068v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as an effective method for enhancing the reasoning capabilities of large language models (LLMs). However, existing methods suffer from insufficient precision in feedback on generated answer trajectories and exh...

📖 Read original article


133. CorePath: A Breast-Specialized Pathology Foundation Model for Core Needle Biopsy Diagnosis and Risk-Controlled Report Generation ​

Author: Ting Yin, Danning Li, Chen Shu, Xiaoxia Yao, Boyu Fu, Yujing Chang, Tianyu Shi, Mengna Feng, Jie Chen, Jing Fu, Xiuli Xiao, Tianlin Li, Mumin Shao, Jiaxin Bi, Wenchuan Zhang, Xiaoyan Wu, Xiao Han, Zhang Zhang, Yuhao Yi, Hong Bu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, stat.AP

arXiv:2608.03079v1 Announce Type: cross Abstract: Breast core needle biopsy (CNB) is central to breast cancer diagnosis yet remains challenging because limited tissue sampling, lesion heterogeneity, and subtle morphologic overlap can obscure subtype distinctions. We developed CorePath, a breast-spec...

📖 Read original article


134. Causal Inference with Unstructured Outcomes ​

Author: Kevin Christian Wibisono, Yixin Wang
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.03085v1 Announce Type: cross Abstract: Causal inference has traditionally centered on scalar outcomes: whether a patient recovers, how much a worker earns, or how many visits a website receives. Modern studies increasingly ask causal questions about outcomes with richer form, such as clin...

📖 Read original article


135. Automatic Patient-Specific Microwave Ablation Planning Accelerated by a Physics-Guided Deep Learning Model ​

Author: Seonaeng Cho, Minjee Seo, Minju Seol, Juil Park, Joon Ho Kwon, Kyungho Yoon
Published: 8/5/2026, 4:00:00 AM
Categories: eess.IV, cs.LG

arXiv:2608.03086v1 Announce Type: cross Abstract: Microwave ablation (MWA) is a promising minimally invasive treatment for liver tumors, but its therapeutic outcome strongly depends on patient-specific planning of antenna insertion trajectory, power, and treatment duration. Accurate numerical simula...

📖 Read original article


136. VIVID: A Culturally Grounded Benchmark Exposing the Figurative Language Gap in Vietnamese NLP ​

Author: Tu Tran Do, Nhat Ngoc Nguyen, Khanh-Tung Tran, Hoang D. Nguyen, Tu Minh Phuong, Long Hoang Dang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.03095v1 Announce Type: cross Abstract: We present VIVID (Vietnamese Idioms for Validation and Interpretation Depth), the first systematic benchmark for evaluating culturally grounded figurative language understanding in Vietnamese. VIVID comprises 1,636 idioms and proverbs annotated with ...

📖 Read original article


137. DP-MemView: A Memory Interface for Attribute-Level Transcript Privacy in Long-Term LLM Agents ​

Author: Jong Wook Kim, Byoungjae Min, Kennedy Edemacu, Yoonhyuk Choi, Sae-Hong Cho, Beakcheol Jang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG

arXiv:2608.03130v1 Announce Type: cross Abstract: Long-term memory enables persistent personalization in LLM agents, but repeated memory-conditioned responses can cumulatively reveal protected attributes even when they are never stated explicitly. We formalize this threat as adaptive transcript priv...

📖 Read original article


138. Minimax-Optimal Semiparametric Contextual Dynamic Pricing with Multimodal Revenue ​

Author: Xueping Gong, Zhuoluo Zhang, Zhaowei Miao, Jiheng Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2608.03142v1 Announce Type: cross Abstract: We study contextual dynamic pricing with arbitrary covariate sequences and bounded, possibly nonbinary purchase quantities. Demand follows a semiparametric surplus-index model with an unknown linear valuation parameter and an unknown H"older-smooth ...

📖 Read original article


139. Surrogate Substitution Preserves PHI Detectability: A Multi-Detector Equivalence Study ​

Author: Qiming Bao, Sherry J. H. Feng, Kim Chester Eugenio, Meng Fon
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.03172v1 Announce Type: cross Abstract: Structure-preserving de-identification replaces protected health information (PHI) with realistic same-type surrogates -- "Anna S." becomes "Maria S.", not [NAME] -- so that clinical text stays fluent and downstream tools keep working. But this only ...

📖 Read original article


140. DRIFT: Derailing Denoising Trajectories of Flow-Matching VLAs with Adversarial Patch Attack ​

Author: Hoseong Tae, Jong-Seok Lee
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.03207v1 Announce Type: cross Abstract: Flow-matching vision-language-action (VLA) models such as pi0 generate robot actions by integrating a learned denoising velocity field, and have been reported to resist adversarial perturbations that readily fool autoregressive VLAs. We show that thi...

📖 Read original article


141. ShielDroid: A Hybrid Approach Integrating Machine and Deep Learning for Android Malware Detection ​

Author: Md Faisal Ahmed, Zarin Tasnim Biash, Abu Raihan Shakil, Ahmed Ann Noor Ryen, Arman Hossain, Faisal Bin Ashraf, Muhammad Iqbal Hossain
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.03250v1 Announce Type: cross Abstract: The rapid advancement of modern technology has led to a significant increase in the use of smart devices, such as smartphones and tablets, resulting in the widespread adoption of mobile applications. Although applications are required to undergo malw...

📖 Read original article


142. Task-Oriented Candidate-Latent Feedback for Coarse-to-Fine Sensing in Distributed OFDM-ISAC Networks ​

Author: Shiv Shankar, Radha Krishna Ganti, J Klutto Milleth
Published: 8/5/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.03319v1 Announce Type: cross Abstract: Future integrated sensing and communication (ISAC) architectures separate the sensing entity (SE) that acquires measurements from the sensing function (SF) that performs inference, creating a need for compact, task-oriented feedback on the SE-SF inte...

📖 Read original article


143. A Direct Route to Markov Chain Convergence via Asymptotic Equivalence with the Target ​

Author: Patrick Forr'e
Published: 8/5/2026, 4:00:00 AM
Categories: math.PR, cs.LG, math.ST, stat.CO, stat.ML, stat.TH

arXiv:2608.03353v1 Announce Type: cross Abstract: For a Markov kernel $T$ with an invariant probability measure $\pi$, we give a self-contained proof of the Markov chain convergence theorem via a criterion called asymptotic equivalence with the target. It assumes two parts about the Lebesgue decompo...

📖 Read original article


144. Conformal risk control for model-form uncertainty in parametric non-intrusive reduced-order models ​

Author: Edgar Jaber (CB, ENS Paris Saclay), R'emy Vallot (CB, Michelin), Thibault Dairay (CB, Michelin), Mathilde Mougeot (CB, ENSIIE, ENS Paris Saclay)
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.03360v1 Announce Type: cross Abstract: Non-intrusive reduced-order models (NIROMs) have become a standard tool for approximating parametric partial differential equations from computer design of experiments while significantly reducing computational costs. However, assessing the reliabili...

📖 Read original article


145. LLM-Derived Priors for Thompson Sampling in Cold-Start Comment Recommendation ​

Author: Eugene Lee, Oseong Choi, Byungsoo Kang, Taeyeong Jang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.03382v1 Announce Type: cross Abstract: Multi-armed bandit algorithms, especially Thompson sampling, are widely used in online recommendation. Despite their ability to adapt from online feedback, these methods often suffer from cold-start limitations when newly introduced arms have little ...

📖 Read original article


146. SRAP: SVD-Refined Adversarial Perturbations for Imperceptible Face-Swap Defense ​

Author: Sungwon Cho, Kwanghyun Ko, Myungjoo Kang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.03395v1 Announce Type: cross Abstract: Deepfake technologies pose increasing threats to facial privacy and identity security, motivating proactive defenses that protect facial images before misuse. Although adversarial perturbations generated by projected gradient descent (PGD) can disrup...

📖 Read original article


147. Distilled Roads: Generalisable Road Network Extraction Across Sensors, Resolutions, and Region ​

Author: Sanayya, Rakshith Sathish, Ashwathi Nambiar
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.03407v1 Announce Type: cross Abstract: Road network segmentation from satellite imagery remains challenging due to large geographic variation in road appearance, occlusions, and domain shifts introduced by differing resolutions and sensors. Existing models, typically trained under narrow ...

📖 Read original article


148. AI World Cup 2026: Benchmarking Large Language Models for End-to-End Football Tournament Prediction ​

Author: Jonaid Shianifar, Iias Faiud
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.03416v1 Announce Type: cross Abstract: Large language models (LLMs) are now regularly asked to forecast real-world events, but comparisons are often difficult because models receive different information, use different tools, and are evaluated under different rules. This paper reports the...

📖 Read original article


149. Dual-domain U-Nets with embedded back projection operators for motion-resolved 4D CBCT reconstruction ​

Author: Ivo Herzig, Pascal Paysan, Daniel Barco, Marc Andr'e Stadelmann, Frank-Peter Schilling, Igor Peterlik, Michal Walczak, Lijin Aryananda, Woo Sang Ahn, Rudolf Marcel F"uchslin, Lukas Lichtensteiger
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.03430v1 Announce Type: cross Abstract: Four-dimensional cone beam CT (4D CBCT) is important for image-guided radiation therapy of thoracic cancers, but its use is limited by long scan times, causing high patient dose and motion/sparse-sampling artifacts. We propose a deep learning method ...

📖 Read original article


150. FedRings: A Scalable and Topology-Aware Federated Learning Framework for LEO Satellite Constellations ​

Author: Ziwu Liu, In^es Pinto Gouveia, Rehana Yasmin, Paulo Esteves-Verissimo, Ali Shoker
Published: 8/5/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.03436v1 Announce Type: cross Abstract: Federated learning over low Earth orbit (LEO) satellite networks is limited by frequent link changes, short contact times, and a highly dynamic topology, making centralized or synchronized training inefficient and hard to scale. To address this, we p...

📖 Read original article


151. Dynamically Allocating Evaluation Effort for Model Ranking ​

Author: Vil'em Zouhar, Julia Kreutzer, Alon Lavie, Tom Kocmi, Matt Post, Ond\v{r}ej Bojar, Mrinmaya Sachan
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.03437v1 Announce Type: cross Abstract: While human evaluation is the gold standard in many NLP tasks, it suffers from prohibitive costs and poor scalability. When identifying top-performing models, typical evaluation protocols waste effort by exhaustively evaluating all models on the enti...

📖 Read original article


152. Quality Control Algorithms for Pattern Counting ​

Author: Cassandra Marcussen, Ronitt Rubinfeld, Madhu Sudan
Published: 8/5/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, math.CO, math.PR

arXiv:2608.03439v1 Announce Type: cross Abstract: In recent work, Marcussen, Rubinfeld, and Sudan introduced the notion of quality control problems, which aim to capture the task of determining if a given input is truly random. Formally, their goal is to accept typical inputs from the specified dist...

📖 Read original article


153. When Correct Solutions Repeat: Rarity-Aware Credit Redistribution for GRPO ​

Author: Zhe Cao, Miaowen Wen, Fangjiong Chen
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.03467v2 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) com- monly optimizes each correct completion as an independent learning signal. In GRPO, this completion-level uniformity creates structure-level skew: recurring correct solution forms accumulate ...

📖 Read original article


154. Should the Boundary Term Be Learned in Reflected Diffusion? Conormal Trace and Reflection Masking ​

Author: Ziyue Wang, Takafumi Kanamori
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.03469v1 Announce Type: cross Abstract: We study score learning for reflected diffusion on bounded domains. Reflection keeps trajectories feasible but does not ensure that the learned score satisfies the boundary behavior implied by the forward process. With implicit score matching, integr...

📖 Read original article


155. Continue or Replan? Bernoulli-Continuation Policy Learning for Adaptive Horizon Execution ​

Author: Weichen Xu, Zhenhua Liu, Lin Luo, Yaobo Liang, Chengtang Yao, Qingyu Mei, Jian Cao, Xixin Cao, Xing Zhang, Jiaolong Yang, Baining Guo
Published: 8/5/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG

arXiv:2608.03483v1 Announce Type: cross Abstract: Existing chunk-based Vision-Language-Action (VLA) models execute a fixed number of actions (i.e., execution horizon) before replanning, turning replanning into a task-agnostic periodic schedule that is independent of task progress. As a result, when ...

📖 Read original article


156. Beyond Initialization Loss: A Systematic Study of Token Embedding Initialization Strategies for LLM Vocabulary Extension ​

Author: Raviraj Joshi, Utkarsh Vaidya, Sanjay Singh Chauhan, Niranjan Wartikar
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.03494v1 Announce Type: cross Abstract: Vocabulary extension is an efficient way to adapt pretrained large language models (LLMs) to new languages, but the initialization of newly added token embeddings can strongly affect continued pre-training (CPT) efficiency. We present a systematic st...

📖 Read original article


157. Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks ​

Author: Christophe D. Hounwanou, John Emeka Eze, Ya'e Ulrich Gaba
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2608.03502v1 Announce Type: cross Abstract: Large Language Models (LLMs) have recently shown strong capabilities in reasoning, planning, and tool-use, enabling new forms of autonomous agents. However, LLM-based agents struggle with long-horizon sequential decision tasks that require precise ac...

📖 Read original article


158. How Many Labels Are Enough? ALDA: Active Learning Deployment Advisor for Medical Image Classification ​

Author: Julia Machnio, Mads Nielsen, Mostafa Mehdipour Ghazi
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.03511v1 Announce Type: cross Abstract: Active learning (AL) promises to reduce the cost of medical imaging projects by lowering the number of clinical labels required. However, practical deployment requires committing to a sampling strategy before the full annotation budget is spent, and ...

📖 Read original article


159. Test-Time Augmentation for Tabular-to-Image Classifiers under Distribution Shifts ​

Author: Malena Loza, Felipe Grijalva, Eva Milara, Luis Bote-Curiel, Francisco J. Lara-Abelenda, David Chushig-Muzo
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.03557v1 Announce Type: cross Abstract: Tabular-to-image methods that convert tabular data into visual representations have emerged as a novel paradigm for leveraging the high performance of deep learning models. Despite their advantages, the robustness of these methods under distribution ...

📖 Read original article


160. Enhancing Tabular Learners with Context-Aware Semantic Embeddings ​

Author: G"unther Schindler, Maximilian Schambach, Johannes H"ohne
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.03565v1 Announce Type: cross Abstract: While modern tabular learners excel at capturing statistical patterns, they frequently operate in a semantic vacuum, treating textual features as discrete symbols, ignoring the rich semantics inherent in feature names or cell entries. We propose CASE...

📖 Read original article


161. Divide-and-Conquer: Towards Generalizable Amortized Bayesian Inference for the Drift Diffusion Model ​

Author: Yufei Wu, Shanqing Gao, Andreas Voss, Francis Tuerlinckx
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP, stat.ME

arXiv:2608.03566v1 Announce Type: cross Abstract: The drift diffusion model (DDM) is a cornerstone of cognitive decision-making research. Although numerous estimation methods exist, researchers continue to seek inference approaches that are both fast and flexible across diverse study designs. Amorti...

📖 Read original article


162. Adversarial Fast-Moving Real-World Domains as Test Beds for Benchmarking AI Scientist Capabilities ​

Author: William Bolton, Philip Torr
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG, cs.MA

arXiv:2608.03569v1 Announce Type: cross Abstract: Benchmarking the ability of AI scientists to generate novel ideas is notoriously difficult. Existing benchmarks in this field have made progress in evaluating scientific reasoning and research replication, but often rely on synthetic tasks or retrosp...

📖 Read original article


163. SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs ​

Author: Kejian Zhu, Zhuoran Jin, Shangqing Tu, Hongbang Yuan, Yushi Bai, Kang Liu, Juanzi Li, Jun Zhao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.03573v1 Announce Type: cross Abstract: Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) exhibit fundamentally different behaviors in enhancing multi-task reasoning for large language models (LLMs). Our preliminary experiments revealed a phenomenon: SFT suffers from severe task...

📖 Read original article


164. FOUND-AF: Benchmarking ECG Foundation Models for Atrial Fibrillation Detection ​

Author: Amirhossein Taleshinosrati, Yangyang Wang, Atitaya Phoemsuk, Vahid Abolghasemi, Naser Hossein Motlagh, Sadasivan Puthusserypady, Daniel Teichmann, Abdolrahman Peimankar
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.03597v1 Announce Type: cross Abstract: Atrial fibrillation (AF) is the most common sustained cardiac arrhythmia and is associated with increased risks of stroke, heart failure, and mortality. Recent ECG foundation models offer transferable representations for automated AF detection. Howev...

📖 Read original article


165. Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents ​

Author: William Bolton, Philip Torr
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.03606v1 Announce Type: cross Abstract: Clinical development is sequential decision-making under uncertainty, where a sponsor must plan a portfolio of experiments from heterogeneous evidence. We study this setting by framing oncology clinical development as an offline decision-making probl...

📖 Read original article


166. Cross-Layer Interaction under Weight-Space Ablation: A Closed-Form Attention Jacobian Bound and a Test on a Real Pretrained Model ​

Author: Abdallah Khemais
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.03629v1 Announce Type: cross Abstract: A companion paper studies when activation patching and weight-space ablation agree, inside an idealized model where a conditional computation is carried additively through a residual stream. For the one composition in that model where two carriers ar...

📖 Read original article


167. Conditionally Identifiable Latent-Environment Modeling for Out-of-Distribution Recommendation ​

Author: Qianqian Wang, Wenwu Gong, Yunshan Li, Zhenqing Wu, Ruili Wang, Lili Yang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.03647v1 Announce Type: cross Abstract: Out-of-distribution (OOD) recommendation is vulnerable to preference shifts induced by a latent environment. Existing methods can infer latent states from logged interactions, yet the statistical meaning of the latent environment and its effect on pr...

📖 Read original article


168. Accelerating Dynamic Graph Clustering on GPU Architectures with cuGraph ​

Author: Nelson Aloysio Reis de Almeida Passos, Emanuele Carlini, Salvatore Trani
Published: 8/5/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.03695v1 Announce Type: cross Abstract: This work addresses community detection in temporal networks through GPU-accelerated extensions of spectral clustering and modularity-based algorithms originally designed for static graphs. Built on the NVIDIA RAPIDS ecosystem, the framework enables ...

📖 Read original article


169. Less Traffic, Better Outcomes: Competition-Aware Request Dispatch in Real-Time Ad Exchanges ​

Author: Jonaid Shianifar, Blaz Mramor, Fangda Zou, Matthieu C. Martin, Xingsheng Guo, Zhihua Zhu, Rong Zhou, Bichen Shi
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.03705v1 Announce Type: cross Abstract: Real-time bidding (RTB) ad exchanges typically forward nearly all incoming requests to demand-side platforms (DSPs), even though only a small fraction receive bids. This over-distribution weakens auction outcomes: DSPs throttle participation under co...

📖 Read original article


170. Attention is Case-Sensitive ​

Author: Maximilian Dillitzer, Tin Stribor Sohn, Jason J. Corso, Michael Auerbach
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2608.03711v1 Announce Type: cross Abstract: In human visual perception, uppercase lettering serves as a natural salience cue that captures attention within lowercase text. In this paper, we present a systematic empirical characterization study revealing that Large Language Models (LLMs) exhibi...

📖 Read original article


171. Can LLMs Test Terminal User Interfaces? ​

Author: Chao Peng, Ruida Hu, Ajitha Rajan, Tegawend'e F Bissyand'e, Jacques Klein, Cuiyun Gao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2608.03743v1 Announce Type: cross Abstract: Terminal User Interfaces (TUIs) combine the stateful, screen-oriented behaviour of GUIs with terminal deployment and are now common in developer tools. Yet they lack a dedicated testing methodology. We survey 197 real-world TUI applications: only 12%...

📖 Read original article


172. Computing Actual Causes for Neural Network Predictions under Structured Causal Inputs ​

Author: Jannick Strobel, Muqsit Azeem, Stefan Leue
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO

arXiv:2608.03772v1 Announce Type: cross Abstract: Explaining the predictions of neural networks is a central challenge in trustworthy AI. Existing explanation methods, such as those based on feature attribution or minimal sufficient sets, typically treat input features as independent, which can yiel...

📖 Read original article


173. Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss ​

Author: Bakbergen Ryskulov, Iker Garc'ia-Ferrero, David Montero, David Jansen, Ali Hashemi, Jezabel R. Garcia, Antonio Tiene, Rom'an Or'us
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.03796v1 Announce Type: cross Abstract: Small language models are often the only option for deployment under tight latency, cost, and on-premises constraints, but they are rarely trained from scratch: a compressed model is usually recovered through knowledge distillation (KD). This recover...

📖 Read original article


174. M-GATE: Multilingual Grammar, Accuracy in Translation, and Efficiency Benchmark for Large Language Models ​

Author: Tom'a\v{s} Burkert, Angelika Peljak-{\L}api'nska, David Zelen'y
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.03803v1 Announce Type: cross Abstract: Multilingual language models are deployed across a hundred or more languages, yet most benchmarks test whether a model can perform a task in a language rather than whether it commands the language itself, conflating fluency with proficiency. We int...

📖 Read original article


175. Geo-Embed: Towards Unified Multimodal Embeddings for Urban Understanding ​

Author: Jiapeng Li, Yong Li, Junjie Zhou, Fan Zhang, Yu Liu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.03826v1 Announce Type: cross Abstract: Geospatial and urban applications increasingly require models to compare heterogeneous evidence across street-view imagery, remote-sensing observations, text descriptions, region proposals, and temporal change cues. However, existing multimodal embed...

📖 Read original article


176. Sensitivity, Causality, and Repair Dissociate: A Layer-Wise Analysis of Perturbation Robustness and Its Scaling ​

Author: Nathan Labiosa, David Buff, Ena Nayak, Erica Donno
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.03842v1 Announce Type: cross Abstract: When a language model fails on surface-perturbed input (typos, OCR noise, homophones), "which layer is responsible" has three natural operationalizations: where representations diverge most (sensitivity), where restoring clean activations recovers th...

📖 Read original article


177. ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities? ​

Author: Tianyi Guan, Yiding Wang, Haotong Yang, Siyuan Cao, Shirui Liu, Yi Hu, Jiaqi Li, Muhan Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.03874v1 Announce Type: cross Abstract: Modern agent frameworks equip large language models with external skill libraries to solve complex tasks. However, it remains unclear whether these systems can effectively evolve their skills and whether the resulting skills improve task-solving capa...

📖 Read original article


178. Socially Grounded Agentic AI: Coordinating Plural Perspectives through Social Theory ​

Author: Matt Ratto, Abhishek Moturu, Daniel Silver
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2608.03910v1 Announce Type: cross Abstract: As AI systems are deployed across increasingly diverse social contexts, alignment can no longer be framed as the optimization of a single, unified set of values. Instead, systems must be able to recognize, represent, and respond to multiple legitimat...

📖 Read original article


179. Robust Low-Tubal-Rank Tensor Completion under Cross-Concentrated Sampling ​

Author: HanQin Cai, Longxiu Huang, Jing Qin, Chengyue Wu
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, cs.NA, math.IT, math.NA

arXiv:2608.03928v2 Announce Type: cross Abstract: Tensor cross-concentrated sampling (t-CCS) bridges entrywise sampling and t-CUR slice-wise sampling by observing entries only within selected horizontal and lateral slices. Existing t-CCS completion methods, however, assume that the observations are ...

📖 Read original article


180. Logic Before Language: Pre-pretraining on Formal Derivations Fosters Skill Acquisition and Compressibility ​

Author: Jo-Ku Cheng, Nikolaos Aletras, Marco Valentino
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.03930v1 Announce Type: cross Abstract: Pre-pretraining language models (LMs) on symbolic data can accelerate and improve natural language acquisition. However, existing pre-pretraining tasks, such as Dyck and procedural algorithms, rely on narrow primitives that fail to capture the expres...

📖 Read original article


181. Information-Geometric Forward Policy Training in GFlowNets ​

Author: Yordan Raykov, Rodrigo Veiga
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.03967v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) have emerged as a flexible framework for amortised inference over discrete and mixed discrete-continuous objects, requiring only an unnormalised target density specified through a reward. In this work, we formulat...

📖 Read original article


182. Representing Random Utility Choice Models with Neural Networks ​

Author: Ali Aouad, Antoine D'esir
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2207.12877v3 Announce Type: replace Abstract: Motivated by the successes of deep learning, we propose a class of neural network-based discrete choice models, called RUMnets, inspired by the random utility maximization (RUM) framework. This model formulates the agents' random utility function u...

📖 Read original article


183. MambaTS: Improved Selective State Space Models for Long-term Time Series Forecasting ​

Author: Xiuding Cai, Xueyao Wang, Yaoyao Zhu, Yu Yao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2405.16440v2 Announce Type: replace Abstract: In recent years, Transformers have become the de-facto architecture for long-term time series forecasting (LTSF), yet they face challenges associated with the self-attention mechanism, including quadratic complexity and permutation-invariant bias. ...

📖 Read original article


184. CollaFuse: Collaborative Diffusion Models ​

Author: Simeon Allmendinger, Domenique Zipperling, Lukas Struppek, Niklas K"uhl
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2406.14429v4 Announce Type: replace Abstract: In the landscape of generative artificial intelligence, diffusion-based models have emerged as a promising method for generating synthetic images. However, the application of diffusion models poses numerous challenges, particularly concerning data ...

📖 Read original article


185. Patient-centered data science: an integrative framework for evaluating and predicting clinical outcomes in the digital health era ​

Author: Mohsen Amoei, Dan Poenaru
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY

arXiv:2408.02677v2 Announce Type: replace Abstract: This study proposes a novel, integrative framework for patient-centered data science in the digital health era. We developed a multidimensional model that combines traditional clinical data with patient-reported outcomes, social determinants of hea...

📖 Read original article


186. Adversarial Purification by Consistency-aware Latent Space Optimization on Data Manifolds ​

Author: Shuhai Zhang, Jiahao Yang, Hui Luo, Jie Chen, Li Wang, Feng Liu, Bo Han, Mingkui Tan
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2412.08394v2 Announce Type: replace Abstract: Deep neural networks (DNNs) are vulnerable to adversarial samples crafted by adding imperceptible perturbations to clean data, potentially leading to incorrect and dangerous predictions. Adversarial purification has been an effective means to impro...

📖 Read original article


187. Rex: A Family of Reversible Exponential (Stochastic) Runge-Kutta Solvers ​

Author: Zander W. Blasingame, Chen Liu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2502.08834v5 Announce Type: replace Abstract: Deep generative models based on neural differential equations have become state-of-the-art for many generation tasks. These models rely on ODE/SDE solvers that integrate from a prior distribution to the data distribution; in many applications it is...

📖 Read original article


188. Virtual Patients, Real Gains: Digital Twin-Based Simulated CT for Multitask Lung Nodule Analysis ​

Author: Fakrul Islam Tushar, Lavsen Dahal, Paul Segars, Joseph Y. Lo
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2502.21187v4 Announce Type: replace Abstract: AI-based lung cancer screening is constrained by scarce, annotated CT data, particularly for rare nodule presentations. We investigate whether physics-based, anatomy-informed simulated CT can improve AI performance across three lung-nodule tasks: d...

📖 Read original article


189. Bi-Lipschitz Ansatz for Anti-Symmetric Functions ​

Author: Nadav Dym, Jianfeng Lu, Matan Mizrachi
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, math.CA, quant-ph

arXiv:2503.04263v2 Announce Type: replace Abstract: Motivated by applications to the simulation of quantum many-body systems by neural networks, researchers have suggested several models which are antisymmetric by construction, and can approximate all antisymmetric functions. However, these works ei...

📖 Read original article


190. Improving Sample Efficiency in Multi-Agent Reinforcement Learning for Simulated Football Games via Exploration ​

Author: Amir Baghi, Jens Sj"olund, Joakim Bergdahl, Linus Gissl'en, Alessandro Sestini
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.MA

arXiv:2503.13077v2 Announce Type: replace Abstract: Multi-agent reinforcement learning has shown promise in learning cooperative behaviors in team-based environments. However, such methods often demand extensive training time, which inhibits their application for game-AI in standard game development...

📖 Read original article


191. When Search Teaches Style: Causal Internalization of Tactical Priors in AlphaZero ​

Author: Ruitong Li, Binjie Guo, Aisheng Mo, Guowei Su, Han Wang, Jie Li, Ru Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2504.14636v3 Announce Type: replace Abstract: AlphaZero is normally evaluated as one agent: a policy-value network fused with Monte Carlo tree search. That fusion hides a causal question. When self-play search is given a useful prior, does the network absorb the induced behavior, or does the b...

📖 Read original article


192. VIBE: Vector Index Benchmark for Embeddings ​

Author: Elias J"a"asaari, Ville Hyv"onen, Matteo Ceccarello, Teemu Roos, Martin Aum"uller
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2505.17810v2 Announce Type: replace Abstract: Approximate nearest neighbor (ANN) search is a performance-critical component of many machine learning pipelines, and rigorous benchmarking is essential for assessing the performance of vector indexes for ANN search. However, the datasets of existi...

📖 Read original article


193. One-Point Contraction: Erasing Representational Separability toward Irreversible Deep Forgetting ​

Author: Jaeheun Jung, Bosung Jung, Suhyun Bae, Donghun Lee
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2507.07754v3 Announce Type: replace Abstract: Machine unlearning is usually evaluated by what the classifier outputs: forget-set accuracy, confidence, membership-inference scores. We show that this is not enough. Across 14 representative unlearning methods on CIFAR-10 and SVHN, a single linear...

📖 Read original article


194. IPPRO: Importance-based Pruning with PRojective Offset for Magnitude-indifferent Structural Pruning ​

Author: Jaeheun Jung, Jaehyuk Lee, Yeajin Lee, Donghun Lee
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2507.14171v3 Announce Type: replace Abstract: Importance-based structured pruning overwhelmingly relies on filter magnitude. This proxy is fundamentally flawed: due to scale invariance, functionally identical filters can receive arbitrarily different importance scores under rescaling. We propo...

📖 Read original article


195. From Generator to Embedder: Harnessing Innate Abilities of Multimodal LLMs via Building Zero-Shot Discriminative Embedding Model ​

Author: Yeong-Joon Ju, Seong-Whan Lee
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR

arXiv:2508.00955v3 Announce Type: replace Abstract: Adapting generative Multimodal Large Language Models (MLLMs) into universal embedding models typically demands resource-intensive contrastive pre-training, while traditional hard negative mining methods suffer from severe false negative contaminati...

📖 Read original article


196. Heteroscedasticity of Denoising Score Matching with Generalised Smooth Noise ​

Author: Juyan Zhang, Rhys Newbury, Xinyang Zhang, Tin Tran, Dana Kulic, Michael Burke
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, stat.AP, stat.ML

arXiv:2508.01597v2 Announce Type: replace Abstract: Score Matching (SM) is a powerful framework for estimating the log-density derivatives of a distribution without calculating its normalizing constants. This capability has made it a cornerstone across multiple domains, from classical sta- tistical ...

📖 Read original article


197. Token Buncher: Shielding LLMs from Harmful Reinforcement Learning Fine-Tuning ​

Author: Weitao Feng, Lixu Wang, Peizhuo Lv, Tianyi Wei, Jie Zhang, Chongyang Gao, Sinong Zhan, Wei Dong
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2508.20697v4 Announce Type: replace Abstract: As large language models (LLMs) continue to grow in capability, so do the risks of harmful misuse through fine-tuning. While most prior studies assume that attackers rely on supervised fine-tuning (SFT) for such misuse, we systematically demonstrat...

📖 Read original article


198. Mechanism of Task-oriented Information Removal in In-context Learning ​

Author: Hakaze Cho, Haolin Yang, Gouki Minegishi, Naoya Inoue
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2509.21012v4 Announce Type: replace Abstract: In-context Learning (ICL) is an emerging few-shot learning paradigm based on modern Language Models (LMs), yet its inner mechanism remains unclear. In this paper, we investigate the mechanism through a novel perspective of information removal. Spec...

📖 Read original article


199. Don't Walk the Line: Boundary Guidance for Filtered Generation ​

Author: Sarah Ball, Andreas Haupt
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2510.11834v3 Announce Type: replace Abstract: Generative models are increasingly paired with safety classifiers that filter harmful or undesirable outputs. A common strategy is to fine-tune the generator to reduce the probability of being filtered, but this can be suboptimal: it often pushes t...

📖 Read original article


200. Enhancing Q-Value Updates in Deep Q-Learning via Successor-State Prediction ​

Author: Lipeng Zu, Hansong Zhou, Xiaonan Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.03836v2 Announce Type: replace Abstract: Deep Q-Networks (DQNs) estimate future returns by learning from transitions sampled from a replay buffer. However, the target updates in DQN often rely on next states generated by actions from past, potentially suboptimal, policy. As a result, thes...

📖 Read original article


201. Target-Aligned Fusion for Decision-Sequence Learning under Dynamics Shift ​

Author: Guojian Wang, Quinson Hon, Xuyang Chen, Lin Zhao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2511.09173v3 Announce Type: replace Abstract: External trajectories can improve offline decision-sequence learning, but dynamics shift may make some source subsequences inconsistent with the target environment. We study how to fuse such trajectories with limited target data for Decision Transf...

📖 Read original article


202. STREAM-VAE: Dual-Path Routing for Slow and Fast Dynamics in Vehicle Telemetry Anomaly Detection ​

Author: Kadir-Kaan "Ozer, Ren'e Ebeling, Markus Enzweiler
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2511.15339v3 Announce Type: replace Abstract: Automotive telemetry data exhibits slow drifts and fast spikes, often within the same sequence, making reliable anomaly detection challenging. Standard reconstruction-based methods, including sequence variational autoencoders (VAEs), use a single l...

📖 Read original article


203. Target-Aware Early Stage Ranking ​

Author: Juhee Hong, Meng Liu, Shengzhi Wang, Jin Zhou, Xiaoheng Mao, Zhao Zhu, Ruochen Liu, Huihui Cheng, Leon Gao, Christopher Leung, Chandra Mouli Sekar, Yijia Liu, Boyang Yu, Tuan Trieu, Dawei Sun, Jeet Kanjani, Rui Li, Jing Qian, Xuan Cao, Minjie Fan, Mingze Gao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.21095v2 Announce Type: replace Abstract: Early Stage Ranking (ESR) in large-scale recommendation systems is dominated by ''user--item decoupling'' Two Tower architectures, which scale efficiently but cannot capture fine-grained, target-aware user--item interactions directly. We propose Ta...

📖 Read original article


204. PRISMA: Improving the Accuracy-Latency Frontier of Diffusion-based PDE Solvers Using Physics-Informed Spectral Attention ​

Author: Medha Sawhney, Abhilash Neog, Mridul Khurana, Anuj Karpatne
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NA, math.NA, stat.ML

arXiv:2512.01370v2 Announce Type: replace Abstract: Diffusion-based solvers for partial differential equations (PDEs) are often bottle-necked by slow gradient-based test-time optimization routines that use PDE residuals for loss guidance. They additionally suffer from optimization instabilities and ...

📖 Read original article


205. PRIVEE: Privacy-Preserving Vertical Federated Learning Against Feature Inference Attacks ​

Author: Sindhuja Madabushi, Haider Ali, Ahmad Faraz Khan, Rui Ning, Hongyi Wu, Chunsheng Xin, Ali. R. Butt, Jin-Hee Cho
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2512.12840v2 Announce Type: replace Abstract: Vertical Federated Learning (VFL) enables collaborative model training across organizations that share common user samples but hold disjoint feature spaces. Despite its potential, VFL is susceptible to feature inference attacks, in which adversaria...

📖 Read original article


206. The Ensemble Schr{\"o}dinger Bridge filter for Nonlinear Data Assimilation ​

Author: Hui Sun
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.18928v4 Announce Type: replace Abstract: This work introduces a novel nonlinear optimal filtering method, termed the Ensemble Schr{"o}dinger Bridge nonlinear filter. The proposed filter combines the standard prediction step with a diffusion-generative-modeling-based analysis step, thereb...

📖 Read original article


207. HERO: Hierarchical Evidential Reasoning Optimization for Radiology Report Generation via Reason-then-Summarize ​

Author: Kun Zhao, Guodong Liu, Hui Ji, Siyuan Dai, Pan Wang, Jifeng Song, Chenghua Lin, Liang Zhan, Haoteng Tang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.03321v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have substantially advanced Radiology Report Generation (RRG), yet aligning them through reinforcement learning (RL) remains challenging due to heterogeneous medical supervision. Vanilla Group Relative Polic...

📖 Read original article


208. When Classes Evolve: A Benchmark and Framework for Stage-Aware Class-Incremental Learning ​

Author: Zheng Zhang, Tao Hu, Xueheng Li, Yang Wang, Rui Li, Jie Zhang, Chengjun Xie
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2602.00573v2 Announce Type: replace Abstract: Class-Incremental Learning (CIL) aims to sequentially learn new classes while mitigating catastrophic forgetting of previously learned knowledge. Conventional CIL approaches implicitly assume that classes are morphologically static, focusing primar...

📖 Read original article


209. On the Limits of Layer Pruning for Generative Reasoning in Large Language Models ​

Author: Safal Shrestha, Anubhav Shrestha, Minwu Kim, Aadim Nepal, Keith Ross
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.01997v3 Announce Type: replace Abstract: Recent work has shown that layer pruning can effectively compress large language models (LLMs) while retaining strong performance on classification benchmarks, often with little or no finetuning. In contrast, generative reasoning tasks, such as GSM...

📖 Read original article


210. When RL Meets Adaptive Speculative Training: A Unified Training-Serving System ​

Author: Junxiong Wang, Fengxiang Bie, Jisen Li, Zhongzhu Zhou, Zelei Shao, Yubo Wang, Yinghui Liu, Qingyang Wu, Avner May, Sri Yanamandra, Ce Zhang, Tri Dao, Percy Liang, Ben Athiwaratkun, Shuaiwen Leon Song, Chenfeng Xu, Xiaoxia Wu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.06932v5 Announce Type: replace Abstract: Speculative decoding can significantly accelerate LLM serving, yet most deployments today disentangle speculator training from serving, treating speculator training as a standalone offline modeling problem. We show that this decoupled formulation i...

📖 Read original article


211. Beyond Solving: Prescriptive Probing for Neural Routing Solvers ​

Author: Reuben Narad, L'eonard Boussioux, Michael Wagner
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.07216v2 Announce Type: replace Abstract: Neural combinatorial optimization (NCO) trains fast heuristics for routing problems, but planners often need more than a single solve: they ask which stop to drop, which transition to preserve, or which subset of stops to remove if a route is infea...

📖 Read original article


212. In-Context Pure Exploration in Continuous Decision Spaces ​

Author: Alessio Russo, Yin-Ching Lee, Ryan Welch, Aldo Pacchiano
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17976v2 Announce Type: replace Abstract: In active sequential testing, also termed pure exploration, a learner is tasked with the goal to adaptively acquire information so as to identify an unknown ground-truth hypothesis with as few queries as possible. This problem has several motivatin...

📖 Read original article


213. SphUnc: Hyperspherical Uncertainty Decomposition and Causal Identification via Information Geometry ​

Author: Rong Fu, Chunlei Meng, Jinshuo Liu, Dianyu Zhao, Yongtai Liu, Yibo Meng, Xiaowen Ma, Wangyu Wu, Yangchen Zeng, Shuaishuai Cao, Simon Fong
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.01168v3 Announce Type: replace Abstract: Reliable decision-making in complex multi-agent systems requires calibrated predictions and interpretable uncertainty. We introduce SphUnc, a unified framework combining hyperspherical representation learning with structural causal modeling. The mo...

📖 Read original article


214. HAPEns: Hardware-Aware Post-Hoc Ensembling for Tabular Data ​

Author: Jannis Maier, Lennart Purucker
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.10582v2 Announce Type: replace Abstract: Ensembling is commonly used in machine learning on tabular data to boost predictive performance and robustness, but larger ensembles often lead to increased hardware demand. We introduce HAPEns, a post-hoc ensembling method that explicitly balances...

📖 Read original article


215. In-Context Molecular Property Prediction with LLMs: A Blinding Study on Memorization and Knowledge Conflicts ​

Author: Matthias Busch, Marius Tacke, Sviatlana V. Lamaka, Mikhail L. Zheludkevich, Christian J. Cyron, Christian Feiler, Roland C. Aydin
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.25857v3 Announce Type: replace Abstract: The capabilities of large language models (LLMs) have expanded beyond natural language processing to scientific prediction tasks, including molecular property prediction. However, their effectiveness in in-context learning remains ambiguous, partic...

📖 Read original article


216. DIB-OD: Preserving the Invariant Core for Robust Heterogeneous Graph Adaptation via Decoupled Information Bottleneck and Online Distillation ​

Author: Yang Yan, Yunxuan Li, Qiuyan Wang, Tianjin Huang, Qiudong Yu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.10882v3 Announce Type: replace Abstract: Graph pre-training can facilitate knowledge transfer across graph datasets, but severe structural and feature shifts may cause negative transfer and adaptation-induced overwriting of reusable knowledge. We propose DIB-OD, a heterogeneous graph adap...

📖 Read original article


217. An Efficient Black-Box Reduction from Online Learning to Multicalibration, and a New Route to $\Phi$-Regret Minimization ​

Author: Gabriele Farina, Juan Carlos Perdomo
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.GT

arXiv:2604.19592v2 Announce Type: replace Abstract: We give a Gordon-Greenwald-Marks (GGM) style black-box reduction from online learning to online multicalibration. Concretely, we show that to achieve high-dimensional multicalibration with respect to a class of functions $\mathcal{H}$, it suffices ...

📖 Read original article


218. Estimating Tail Risks in Language Model Output Distributions ​

Author: Rico Angell, Raghav Singhal, Zachary Horvitz, Zhou Yu, Rajesh Ranganath, Kathleen McKeown, He He
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.22167v3 Announce Type: replace Abstract: Language models are increasingly capable and are being rapidly deployed on a population-level scale. As a result, the safety of these models is increasingly high-stakes. Fortunately, advances in alignment have significantly reduced the likelihood o...

📖 Read original article


219. NPMixer: Hierarchical Neighboring Patch Mixing for Time Series Forecasting ​

Author: Jung Min Choi, Vijaya Krishna Yalavarthi, Lars Schmidt-Thieme
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.07476v2 Announce Type: replace Abstract: Multivariate time series forecasting remains a challenge due to the complexity of local temporal dynamics and global dependencies across multiple variables. In this paper, we propose \textbf{N}eighboring \textbf{P}atching \textbf{Mixer} (\textbf{NP...

📖 Read original article


220. Improving Reproducibility in Evaluation through Multi-Level Annotator Modeling ​

Author: Deepak Pandita, Flip Korn, Chris Welty, Christopher M. Homan
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.13801v2 Announce Type: replace Abstract: As generative AI models such as large language models (LLMs) become more pervasive, ensuring the safety, robustness, and overall trustworthiness of these systems is paramount. However, AI is currently facing a reproducibility crisis driven by unrel...

📖 Read original article


221. A Deployment Audit of Release-Side Risk in Conformal Triage under Prevalence Shift ​

Author: Chengze Li, Xiao Liu, Hanrong Zhang, Haiyang Peng, Yanghao Ruan, Huanhuan Ma, Chunyu Miao, Qichao Zhou, Xiangrong Qi, Philip Yu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2605.20956v2 Announce Type: replace Abstract: Conformal triage converts predictive scores into deployment actions that either release a case, flag it for urgent attention, or defer it to human review. Under an observed change in target-event prevalence, however, marginal coverage and human-rev...

📖 Read original article


222. Sign-Aware Gated Sparse Autoencoders: Modeling Anticorrelated Features with Bi-Jump-ReLU Activations ​

Author: Bartosz Wieciech, Zmnako Awrahman, Marcin Czelej, Victor Hugo Jaramillo Velasquez, Wioletta Stobieniecka
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.28149v3 Announce Type: replace Abstract: Sparse Autoencoders (SAEs) extract interpretable features from Large Language Model activations, but standard variants enforce non-negative latents, so a bidirectional semantic axis (e.g., "pressure too high" vs. "pressure too low") must be split a...

📖 Read original article


223. E4GEN: Event-level Explainable Extreme-Enhanced Time-series Generation ​

Author: Lin Jiang, Dahai Yu, Ximiao Li, Guang Wang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.01634v2 Announce Type: replace Abstract: Generating realistic time series is essential for scientific research and real-world applications. However, existing methods often emphasize overall distributional fidelity while failing to faithfully capture extreme events. To advance existing res...

📖 Read original article


224. FLARE: Diffusion for Hybrid Language Model ​

Author: Yuchen Zhu, Jing Shi, Chongjian Ge, Hao Tan, Yiran Xu, Wanrong Zhu, Jason Kuen, Koustava Goswami, Rajiv Jain, Yongxin Chen, Molei Tao, Jiuxiang Gu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.01774v2 Announce Type: replace Abstract: Autoregressive (AR) large language models (LLMs) have achieved broad practical success, but sequential decoding remains a key bottleneck for low-latency deployment. Recent efficient-inference work has progressed along two axes: reducing the cost of...

📖 Read original article


225. CaliDist: Calibrating Large Language Models via Behavioral Robustness to Distraction ​

Author: Mohammad Anas Jawad, Cornelia Caragea
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2606.05799v2 Announce Type: replace Abstract: Existing calibration methods for Large Language Models (LLMs) often overlook a critical dimension of trustworthiness: a model's behavioral robustness to irrelevant or misleading information. In this paper, we argue that a model's true confidence sh...

📖 Read original article


226. When Behavioral Safety Evaluation Fails: A Representation-Level Perspective ​

Author: Enyi Jiang, Anders Gj{\o}lbye, Yibo Jacky Zhang, Sanmi Koyejo
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2606.08044v2 Announce Type: replace Abstract: Safety evaluation of large language models (LLMs) is largely behavioral: a model is certified safe when it refuses harmful requests and answers benign ones. But refusing on the prompts an auditor happens to try does not show that the model is far f...

📖 Read original article


227. GENERIC-FNO: Embedding Energy Conservation and Entropy Production into Fourier Neural Operators ​

Author: Jason Sulskis, Sathya Ravi
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.08343v4 Announce Type: replace Abstract: We introduce GENERIC-FNO, the first neural operator to embed the full GENERIC (metriplectic) structure of nonequilibrium thermodynamics -- reversible, energy-conserving dynamics and irreversible, entropy-producing dynamics coupled through the degen...

📖 Read original article


228. When Context Returns: Toward Robust Internalization in On-Policy Distillation ​

Author: Xun Wang, Ruishuo Chen, Zhuoran Li, Yu Chen, Longbo Huang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.11627v2 Announce Type: replace Abstract: Recent work has shown that on-policy distillation can internalize privileged context, such as system prompts or task hints, into a student model so that the context is no longer needed at inference time. However, we identify a counterintuitive and ...

📖 Read original article


229. Online Shift Detection and Conformal Adaptation for Deployed Safety Classifiers ​

Author: Jun Wen Leong
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, stat.ML

arXiv:2606.11949v3 Announce Type: replace Abstract: Reasoning models deployed as safety monitors exhibit a systematic vulnerability: reasoning-token budget starvation. Adversarial inputs require $3.3\times$ more reasoning tokens than benign inputs to produce valid safety scores ($T_{50,\text{adv}}{=...

📖 Read original article


230. MoECa: Aligning Feature Reuse with Expert Decomposition in Diffusion Transformers ​

Author: Maoliang Li, Haojing Chen, Jiayu Chen, Zihao Zheng, Xinhao Sun, Hailong Zou, Xiang Chen
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2606.15615v2 Announce Type: replace Abstract: Diffusion Transformers with Mixture-of-Experts (DiT-MoE) improve model capacity under sparse activation, but diffusion inference is still bottlenecked by redundant computation across timesteps. Existing caching methods mainly operate at the token l...

📖 Read original article


231. Symplectic Neural Networks for Learning Non-Separable Hamiltonians ​

Author: Harsh Choudhary, Vyacheslav Kungurtsev, Chandan Gupta, Melvin Leok, Georgios Korpas
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.27029v2 Announce Type: replace Abstract: Hamiltonian Neural Networks (HNNs) integrate physical priors into neural models by learning a system's Hamiltonian, improving generalization and sample efficiency. Identifying the system Hamiltonian from noisy observations of state variables is a c...

📖 Read original article


232. AdaBoosting Text Prompts for Vision-Language Models ​

Author: Seokhee Jin, Changhwan Sung, Sunung Mun, Hoyoung Kim, Jungseul Ok
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.00684v3 Announce Type: replace Abstract: The classification accuracy of pretrained Vision-Language Models (VLMs) relies on the quality of the text prompts. Handcrafted templates and Large Language Model (LLM)-generated descriptions not only make predictions more interpretable, but also en...

📖 Read original article


233. Foundations of Equivariant Deep Learning: Unifying Graph and Sheaf Neural Networks ​

Author: Yoshihiro Maruyama
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.03798v3 Announce Type: replace Abstract: Symmetry is everywhere in nature and society. Geometric deep learning builds architectures respecting group symmetries, whereas topological deep learning organizes computation through cells, incidence relations, and local-to-global structure. In th...

📖 Read original article


234. x-Prediction Is All You Need:Training-Free Accelerated Generation via Endpoint Decodability ​

Author: Xin Peng, Ang Gao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.06114v4 Announce Type: replace Abstract: Diffusion and flow matching models generate high-quality samples, but their ODE samplers often need tens to hundreds of neural function evaluations (NFEs). This remains a practical challenge for released checkpoints, since many accelerators require...

📖 Read original article


235. Image classification via a quantum-inspired strategy involving a mixture of experts ​

Author: Kumari Jyoti, Rohith Babu, Apoorva D. Patel
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, quant-ph

arXiv:2607.07754v2 Announce Type: replace Abstract: Pattern recognition problems arise in a variety of physical image processing situations, and convolutional neural networks are a popular scheme for the required feature extraction and classification tasks. The classical networks use diffusion-based...

📖 Read original article


236. Kernel weighted importance sampling for off-policy evaluation in contextual bandits ​

Author: Joshua Spear, Matthieu Komorowski, Rebecca Pope, Erica E. M. Moodie
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.15067v2 Announce Type: replace Abstract: This article presents a novel estimator for performing off-policy evaluation using only offline data for contextual bandits. The proposed estimator, Kernel-WIS is demonstrated to be asymptotically consistent and to empirically outperform strong bas...

📖 Read original article


237. The Dark Room in the Reward Channel: Dense Prediction Rewards Collapse GRPO-Trained LLM Agents -- and The Channel, Not the Content, Decides What Works ​

Author: Yu Wang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21273v2 Announce Type: replace Abstract: Dense per-step supervision is the standard remedy for sparse-reward long-horizon LLM agents: reward the policy for predicting its next observation, which looks provably safe under potential-based shaping. Published prediction-reward and auxiliary-l...

📖 Read original article


238. Toward Goal-Agnostic Joint-Embedding Predictive Control of Partial Differential Equations ​

Author: Jonathan Gallagher, Roberto Guglielmi
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2607.21644v2 Announce Type: replace Abstract: We present a goal-agnostic control framework for partial differential equations (PDEs) built around an end-to-end joint-embedding predictive architecture (JEPA). A lightweight 2D vision-transformer (ViT) and action-conditioned latent dynamics are t...

📖 Read original article


239. TriGlue: a Biology-Inspired Generative Model for Generating Molecular Glue-Induced Ternary Complex ​

Author: Yuliang Yan, Shuo Yan, Haochun Tang, Yiqin Sun, Enyan Dai
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.22143v2 Announce Type: replace Abstract: Molecular glue degraders have emerged as a promising strategy for targeted protein degradation by inducing ternary complex formation between an E3 ubiquitin ligase and a target protein. Despite their therapeutic potential, computational design of m...

📖 Read original article


240. DHRCL:Training Code LLMs with Dense Hierarchical Rewards and Curriculum Learning ​

Author: Shuhang Wang, Ziming Li, Hui Cheng
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.26457v2 Announce Type: replace Abstract: Reinforcement learning is a natural post-training paradigm for code-oriented large language models because generated programs can be evaluated through parsing, execution, unit tests, and structural analysis. However, existing methods often rely on ...

📖 Read original article


241. Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? ​

Author: Perry Dong, Ron Polonsky, Dorsa Sadigh, Chelsea Finn
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.27203v2 Announce Type: replace Abstract: Pre-training followed by fine-tuning has become the dominant recipe for learning performant policies, and in value-based reinforcement learning (RL) this raises a natural question: given a pretrained policy, should the Q-function be pretrained on o...

📖 Read original article


242. Paris as a 15-Minute City: An Explainable AI Perspective ​

Author: Andr'as J. Moln'ar, Csaba I. Sidl'o, Rita R'onai, Domonkos R'ozsay
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.00815v2 Announce Type: replace Abstract: The 15-minute city promotes access to everyday services within a short walk or bicycle ride, but its relationship with observed mobility remains difficult to quantify. We investigate this relationship in the Paris metropolitan area using mobility t...

📖 Read original article


243. EulerLoRA: Rank-Driven Jump Dynamics for Calibrated Parameter-Efficient Fine-Tuning ​

Author: Srinivas Anumasa, Dianbo Liu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.01142v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning, but standard LoRA produces a single deterministic model and does not directly support predictive uncertainty estimation. We introduce EulerLoRA, a stochastic extension of LoRA that...

📖 Read original article


244. Beckmann Transport Models: From Autonomous Flows to One-Step Maps ​

Author: Lee Cheuk-Kit, Florentin Coeurdoux, Peter Potaptchik, Yilun Du, Michael Samuel Albergo, Eric Vanden-Eijnden
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.01692v2 Announce Type: replace Abstract: We propose an instantiation of flow matching that relies on a time-independent velocity field (an \emph{autonomous flow}) to exactly map between two distributions, so long as the target is singular, i.e.\ supported on a lower-dimensional data manif...

📖 Read original article


245. LEAP: Lean Environment-Feedback via Adaptive Pruning for Code RL in GPU Kernel Generation ​

Author: Tankun Li, Zhi Chen, Yaohua Tang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.01804v2 Announce Type: replace Abstract: Post-training large language models (LLMs) via reinforcement learning (RL) has significantly advanced code generation capabilities. To bypass the heavy memory footprint of critic networks, current state-of-the-art frameworks leverage critic-free pa...

📖 Read original article


246. RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States ​

Author: Yi Yang, Zhennan Chen, Yihong Zhuang, Tiehan Fan, Yinan Chen, Jian Li, Jian Yang, Ying Tai
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.02508v2 Announce Type: replace Abstract: Learning-based memory systems for self-evolving LLM agents face two tightly coupled challenges. First, trajectory-indexed utilities grow with the interaction history, thereby dispersing limited feedback over an ever-expanding state space. Second, b...

📖 Read original article


247. Convergence analysis of controlled particle systems arising in deep learning: from finite to infinite sample size ​

Author: Huafu Liao, Alp'ar R. M'esz'aros, Chenchen Mou, Chao Zhou
Published: 8/5/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.PR, stat.ML

arXiv:2404.05185v4 Announce Type: replace-cross Abstract: This paper deals with a class of neural SDEs and studies the limiting behavior of the associated sampled optimal control problems as the sample size grows to infinity. The neural SDEs with $N$ samples can be linked to the $N$-particle systems...

📖 Read original article


248. Robust Biharmonic Skinning Using Geometric Fields ​

Author: Ana Dodik, Vincent Sitzmann, Justin Solomon, Oded Stein
Published: 8/5/2026, 4:00:00 AM
Categories: cs.GR, cs.LG

arXiv:2406.00238v3 Announce Type: replace-cross Abstract: Bounded bihramonic weights are a popular tool used to rig and deform characters for animation, to compute reduced-order simulations, and to define feature descriptors for geometry processing. They necessitate tetrahedralizing the volume bound...

📖 Read original article


249. Efficient unsupervised domain adaptation via self-supervised vision transformer and synergistic cross-domain alignment ​

Author: Ali Abedi, Q. M. Jonathan Wu, Ning Zhang, Farhad Pourpanah
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2407.21311v2 Announce Type: replace-cross Abstract: Unsupervised domain adaptation (UDA) aims to mitigate domain shift, where the distribution of labeled source data differs from that of unlabeled target data. Despite recent advances, existing methods often rely on fine-tuning large backbone m...

📖 Read original article


250. Efficient quantum-enhanced classical simulation for patches of quantum landscapes ​

Author: Sacha Lerch, Ricard Puig, Manuel S. Rudolph, Armando Angrisani, Tyson Jones, M. Cerezo, Supanut Thanasilp, Zo"e Holmes
Published: 8/5/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, stat.ML

arXiv:2411.19896v2 Announce Type: replace-cross Abstract: Understanding the capabilities of classical simulation methods is key to identifying where quantum computers are advantageous. Not only does this ensure that quantum computers are used only where necessary, but also one can potentially identi...

📖 Read original article


251. Prediction-Enhanced Monte Carlo: A Machine Learning View on Control Variate ​

Author: Fengpei Li, Haoxian Chen, Jiahe Lin, Arkin Gupta, Xiaowei Tan, Honglei Zhao, Gang Xu, Yuriy Nevmyvaka, Agostino Capponi, Henry Lam
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.CE, cs.LG, q-fin.PR

arXiv:2412.11257v4 Announce Type: replace-cross Abstract: For many complex simulation tasks spanning areas such as healthcare, engineering, and finance, Monte Carlo (MC) methods are invaluable due to their unbiased estimates and precise error quantification. Nevertheless, Monte Carlo simulations oft...

📖 Read original article


252. Strong bounds for large-scale Minimum Sum-of-Squares Clustering ​

Author: Anna Livia Croella, Veronica Piccialli, Antonio M. Sudoso
Published: 8/5/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2502.08397v3 Announce Type: replace-cross Abstract: Clustering is a fundamental technique in data analysis and machine learning, used to group similar data points together. Among various clustering methods, the Minimum Sum-of-Squares Clustering (MSSC) is one of the most widely used. MSSC aims ...

📖 Read original article


253. Beyond Either-Or Reasoning: Transduction and Induction as Cooperative Problem-Solving Paradigms ​

Author: Janis Zenkner, Tobias Sesterhenn, Christian Bartelt
Published: 8/5/2026, 4:00:00 AM
Categories: cs.PL, cs.AI, cs.LG

arXiv:2505.14744v3 Announce Type: replace-cross Abstract: Traditionally, in Programming-by-example (PBE) the goal is to synthesize a program from a small set of input-output examples. Lately, PBE has gained traction as a few-shot reasoning benchmark, relaxing the requirement to produce a program art...

📖 Read original article


254. Uncovering Spontaneous Physics Representations in In-Context Learning ​

Author: Yeongwoo Song, Jaeyong Bae, Dong-Kyum Kim, Hawoong Jeong
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2508.12448v2 Announce Type: replace-cross Abstract: In-context learning (ICL) lets large language models (LLMs) solve new tasks from prompts alone, across an ever-widening range of domains, yet the mechanisms underlying this ability remain poorly understood. Physical systems offer a controlled...

📖 Read original article


255. SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA ​

Author: Haozhou Xu, Dongxia Wu, Matteo Chinazzi, Ruijia Niu, Rose Yu, Yi-An Ma
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2509.25459v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) show promise in generating long-form scientific explanations that synthesize evidence and connect multiple factors. However, in long-form scientific question answering, LLMs often hallucinate, producing unsupporte...

📖 Read original article


256. Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain ​

Author: L'eo Boisvert, Abhay Puri, Chandra Kiran Reddy Evuru, Nazanin Sepahvand, Nicolas Chapados, Quentin Cappart, Jason Stanley, Alexandre Lacoste, Krishnamurthy Dj Dvijotham, Alexandre Drouin
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2510.05159v5 Announce Type: replace-cross Abstract: While finetuning AI agents on interaction data -- such as web browsing or tool use -- improves their capabilities, it also introduces critical security vulnerabilities within the agentic AI supply chain. We show that adversaries can effective...

📖 Read original article


257. Prescribed-Basis Coefficient-to-Coefficient Neural Operator for Partial Differential Equations ​

Author: Chuqi Chen, Yang Xiang, Weihong Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA

arXiv:2510.10350v3 Announce Type: replace-cross Abstract: Operator learning provides a data-driven approach to approximating solution operators of partial differential equations, but its effectiveness depends strongly on how input and output functions are represented. Point-value representations can...

📖 Read original article


258. Bridging Prediction and Attribution: Identifying Forward and Backward Causal Influence Ranges Using Assimilative Causal Inference ​

Author: Marios Andreou, Nan Chen
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, physics.data-an, stat.ME, stat.TH

arXiv:2510.21889v2 Announce Type: replace-cross Abstract: Causal inference identifies cause-and-effect relationships between variables. While traditional approaches rely on data to reveal causal links, a recently developed method, assimilative causal inference (ACI), integrates observations with dyn...

📖 Read original article


259. Design Criteria for SGD Preconditioners: Local Conditioning, Noise Floors, and Basin Stability ​

Author: Mitchell Scott, Tianshi Xu, Ziyuan Tang, Alexandra Pichette-Emmons, Qiang Ye, Yousef Saad, Yuanzhe Xi
Published: 8/5/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA

arXiv:2511.19716v3 Announce Type: replace-cross Abstract: Stochastic Gradient Descent (SGD) often slows in the late stage of training due to anisotropic curvature and gradient noise. We analyze preconditioned SGD in the geometry induced by a symmetric positive definite matrix $\mathbf{M}$, deriving ...

📖 Read original article


260. AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection ​

Author: Wachiraphan Charoenwet, Kla Tantithamthavorn, Patanamon Thongtanunam, Hong Yi Lin, Minwoo Jeong, Ming Wu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG, cs.SE

arXiv:2601.19138v2 Announce Type: replace-cross Abstract: Secure code review is critical during pre-integration, where Atlassian developers rely on lightweight analysis tools, while deep security assessment is deferred to later stages, delaying feedback and increasing remediation costs. Existing sta...

📖 Read original article


261. Cross-Country Learning for National Infectious Disease Forecasting Using European Data ​

Author: Zacharias Komodromos, Kleanthis Malialis, Artemis Kontou, Panayiotis Kolios
Published: 8/5/2026, 4:00:00 AM
Categories: q-bio.PE, cs.LG

arXiv:2601.20771v2 Announce Type: replace-cross Abstract: Accurate forecasting of infectious disease incidence is critical for public health planning and timely intervention. While most data-driven forecasting approaches rely primarily on historical data from a single country, such data are often li...

📖 Read original article


262. The Signal Horizon: Local Blindness and the Contraction of Pauli-Weight Spectra in Noisy Quantum Encodings ​

Author: Ait Haddou Marwan
Published: 8/5/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2602.14735v2 Announce Type: replace-cross Abstract: The performance of quantum classifiers is typically analyzed through global state distinguishability or the trainability of variational models. This study investigates how much class information remains accessible under locality-constrained m...

📖 Read original article


263. LoBoost: Fast Model-Native Local Conformal Prediction for Gradient-Boosted Trees ​

Author: Vagner Santos, Victor Coscrato, Luben Cabezas, Rafael Izbicki, Thiago Ramos
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2602.22432v2 Announce Type: replace-cross Abstract: Gradient-boosted decision trees are among the strongest off-the-shelf predictors for tabular regression, but point predictions alone do not quantify uncertainty. Conformal prediction provides distribution-free marginal coverage, yet standard ...

📖 Read original article


264. Modeling Matches as Language: A Generative Transformer Approach for Counterfactual Player Valuation in Football ​

Author: Miru Hong, Minho Lee, Geonhee Jo, Hyeokje Cho, Hyunsung Kim, Pascal Bauer, Sang-Ki Ko
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2603.15212v2 Announce Type: replace-cross Abstract: Evaluating football player transfers is challenging because player actions depend strongly on tactical systems, teammates, and match context. Despite this complexity, recruitment decisions often rely on static statistics and subjective expert...

📖 Read original article


265. Towards Interpretable Foundation Models for Retinal Fundus Images ​

Author: Samuel Ofosu Mensah, Camila Roa, Kerol Djoumessi, Philipp Berens
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, stat.CO

arXiv:2603.18846v4 Announce Type: replace-cross Abstract: Foundation models are used to extract transferable representations from large amounts of unlabeled data, typically via self-supervised learning (SSL). However, many of these models rely on architectures that offer limited interpretability, a ...

📖 Read original article


266. HUKUKBERT: Domain-Specific Language Model for Turkish Law ​

Author: Mehmet Utku "Ozt"urk, Tansu T"urko\u{g}lu, Buse Buz-Yalug
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2604.04790v2 Announce Type: replace-cross Abstract: Natural language processing (NLP) advances have powered a generation of LegalTech systems, but Turkish law remains under-served by domain-specific data and models. English has legal encoders such as LEGAL-BERT; no comparable high-volume Turki...

📖 Read original article


267. The Illusion of Superposition? A Principled Analysis of Latent Thinking in Language Models ​

Author: Michael Rizvi-Martel, Guillaume Rabusseau, Marius Mosbach
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2604.06374v2 Announce Type: replace-cross Abstract: Latent reasoning via continuous chain-of-thoughts (Latent CoT) has emerged as a promising alternative to discrete CoT reasoning. Operating in continuous space increases expressivity and has been hypothesized to enable superposition: the abili...

📖 Read original article


268. Variational Approximated Restricted Maximum Likelihood Estimation for Spatial Data ​

Author: Debjoy Thakur
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP

arXiv:2604.07635v2 Announce Type: replace-cross Abstract: This research considers a scalable inference for spatial data modeled through Gaussian intrinsic conditional autoregressive (ICAR) structures. The classical estimation method, restricted maximum likelihood (REML), requires repeated inversion ...

📖 Read original article


269. Injection-Execution Dissociation: A Mechanistic Evaluation of Persistent Memory Attacks and Defenses in Stateful LLM Agents ​

Author: Jun Wen Leong
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2605.08442v5 Announce Type: replace-cross Abstract: We discover that prompt-injection success and tool-execution success are separable safety properties: defenses that block injection do not necessarily block execution, and vice versa. We call this the injection-execution dissociation. In LLM ...

📖 Read original article


270. Conformal Anomaly Detection in Python: Moving Beyond Heuristic Thresholds with nonconform ​

Author: Oliver Hennh"ofer, Maximilian Kirsch, Christine Preisach
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO

arXiv:2605.13642v3 Announce Type: replace-cross Abstract: Most anomaly detection systems output scores rather than calibrated decisions, leaving practitioners to choose thresholds heuristically and without clear statistical interpretation. Conformal anomaly detection addresses this limitation by con...

📖 Read original article


Author: Jialin Lu, Soonho Kong, Rodrigo Stehling, Kaiyu Yang, Zhangyang Wang, Weiran Sun, Wuyang Chen
Published: 8/5/2026, 4:00:00 AM
Categories: cs.LO, cs.AI, cs.CL, cs.LG, cs.SE

arXiv:2605.20244v2 Announce Type: replace-cross Abstract: We present Lean Refactor, a plug-and-play retrieval-augmented agentic framework for multi-objective, controllable, and version-robust refactoring of Lean proofs. LLM-generated proofs are notoriously correct-but-verbose and brittle across libr...

📖 Read original article


272. MedCRP-CL: Continual Medical Image Segmentation via Bayesian Nonparametric Semantic Modality Discovery ​

Author: Ziyuan Gao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2605.20297v2 Announce Type: replace-cross Abstract: Medical image segmentation faces a fundamental challenge in continual learning: data arrives sequentially from heterogeneous sources, yet effective continual learning requires discovering which tasks share sufficient structure to benefit from...

📖 Read original article


273. Algorithms with Polynomially-Improved Approximation Factors for the $2 \rightarrow q$ Norm, and Applications ​

Author: Samuel B. Hopkins, Stefan Tiegel
Published: 8/5/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, math.ST, stat.ML, stat.TH

arXiv:2605.25303v3 Announce Type: replace-cross Abstract: The $2 \rightarrow q$ norm of a matrix $X \in \mathbb{R}^{n \times d}$ is defined as $\lVert X \rVert_{2 \rightarrow q} = \sup_{\lVert v \rVert_2 = 1} \lVert Xv \rVert_q$. We give polynomial-time multiplicative approximation algorithms for th...

📖 Read original article


274. Learning to Translate from Soft to Hard LLM Prompts ​

Author: Pitipat Kongsomjit, Suryansh Goyal, Jacob Whitehill
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2605.27642v2 Announce Type: replace-cross Abstract: Soft prompting, also known as continuous prompting, is a parameter-efficient method for tuning LLMs to specific tasks. Like other machine learning techniques, its parameters encode some hidden procedure: is it possible to train a model to dec...

📖 Read original article


275. Speculative Decoding and the Curse of Multilinguality ​

Author: Nirajan Paudel, Michael Ginn, Luc De Nardi, Alexis Palmer
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2605.30580v2 Announce Type: replace-cross Abstract: Speculative decoding is a popular technique for large language model (LLM) inference, enabling faster generation by drafting multiple tokens with a smaller draft model. However, the effectiveness of speculative decoding has mainly been studie...

📖 Read original article


276. Where Did It Go Wrong? Process-Level Evaluation of Web Agents with Semantic State Tracking ​

Author: Jiwan Chung, JiHyuk Byun, Vibhav Vineet, Seon Joo Kim
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.15673v2 Announce Type: replace-cross Abstract: Web agents act through long interaction sequences, yet existing benchmarks evaluate only terminal success, discarding all process information and offering little guidance on improvement. In this work, we conduct a process-level analysis of we...

📖 Read original article


277. Adversarial observations in probabilistic State-Space Models for robust Reinforcement Learning ​

Author: M. Santos-Pascual, D. R'ios Insua
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2606.20880v2 Announce Type: replace-cross Abstract: Decision-making under partial or adversarial observability requires accurate inference of the environment's latent state and its associated uncertainty. This work analyses adversarial attacks on linear state-space models, where the attacker a...

📖 Read original article


278. Accelerated and Stable Convergence with Anchored Generalized Optimistic Method ​

Author: Motahareh Sohrabi, Jianxin You, Simon Lacoste-Julien, Eduard Gorbunov, Gauthier Gidel
Published: 8/5/2026, 4:00:00 AM
Categories: math.OC, cs.GT, cs.LG

arXiv:2606.21528v2 Announce Type: replace-cross Abstract: We study first-order methods for solving monotone variational inequalities arising in min-max optimization. Classical approaches such as the extragradient method rely on two gradient queries per iteration, which limits their analysis and appl...

📖 Read original article


279. Signed Evidence Flow: Conflict-Aware and Stability-Calibrated Data Analysis ​

Author: Jeffery Opoku, David Banahene
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2606.21875v3 Announce Type: replace-cross Abstract: Modern data analysis usually gives a prediction without showing whether the evidence behind it is clear, conflicting, or stable. Two cases can have the same fitted confidence even when one has mostly agreeing evidence and the other has strong...

📖 Read original article


280. Transplanting, inverting, and preventing a misalignment persona: method-conditional emergent misalignment in Qwen2.5 ​

Author: Lyndon Drake (University of Oxford), Zandi Eberstadt (University of Oxford)
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG

arXiv:2607.04510v2 Announce Type: replace-cross Abstract: Emergent misalignment (EM) --- the broad misbehaviour a language model acquires after fine-tuning on narrow harmful data --- is mediated in Qwen2.5 models by a latent persona direction, and that direction is causal in open weights. Transplant...

📖 Read original article


281. DriftWorld: Fast World Modeling through Drifting ​

Author: Susie Lu, Haonan Chen, Weirui Ye, Yilun Du
Published: 8/5/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2607.15065v2 Announce Type: replace-cross Abstract: Predictive world models enable robots to plan by imagining the outcomes of their actions, but their value for control hinges on generating many rollouts quickly. This creates a bottleneck for diffusion-based world models: multistep sampling m...

📖 Read original article


282. Subjective Risk Decomposition: A New View for Uncertainty Quantification ​

Author: Raghad Alamri, Michele Caprio, Gavin Brown
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2607.15196v2 Announce Type: replace-cross Abstract: We present a novel viewpoint for uncertainty quantification. Uncertainty measures are not primitives, in need of axioms and argumentation, but instead consequences, of higher-level modelling decisions. We show how epistemic and aleatoric unce...

📖 Read original article


283. Fretiq: Browser-Native Electric Guitar String Classification via Engineered Spectral Features and Held-Out Free-Play Evaluation ​

Author: Aadi Garg
Published: 8/5/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS

arXiv:2607.18303v2 Announce Type: replace-cross Abstract: Identifying which string produces a given pitch in monophonic electric guitar audio is a classification challenge: a single pitch can often be produced on multiple strings, with timbral differences largely imperceptible to untrained humans. W...

📖 Read original article


284. 1-Lipschitz Neural Networks on Hadamard Manifolds ​

Author: Davide Murari, Marta Ghirardelli, Ben Adcock, Elena Celledoni, Brynjulf Owren, Carola-Bibiane Sch"onlieb
Published: 8/5/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA

arXiv:2607.19335v3 Announce Type: replace-cross Abstract: Controlling the Lipschitz constant of a neural network is a standard way to promote robustness and stability. Most existing constraining strategies are designed for Euclidean spaces. In this work, we construct and analyze a class of 1-Lipschi...

📖 Read original article


285. Making Single-Cell Data Distillation Auditable: Traceable Real-Cell Coresets via Discrete Min--Max Selection ​

Author: Yaodi Luo, Peize He, Lingbei Meng, Bowen Han, Zheng Lu, Jianqing Zhu, Lian Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: q-bio.GN, cs.AI, cs.LG

arXiv:2607.19426v2 Announce Type: replace-cross Abstract: Large single-cell datasets are expensive to store, curate, and repeatedly reuse for model training. Data distillation can reduce this burden by building smaller training sets. However, many existing methods rely on synthetic cells. These synt...

📖 Read original article


286. CausalForge: A Formally Grounded, Self-Improving Agentic Framework for Automated Research in Causal Inference ​

Author: Jiyuan Tan, Vasilis Syrgkanis
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, econ.EM

arXiv:2607.22511v2 Announce Type: replace-cross Abstract: Automating theoretical research is constrained not only by the generation of candidate results, but also by their reliable evaluation. A common approach is to close the research loop with a large language model (LLM) reviewer. However, such r...

📖 Read original article


287. Cortex: Compact Behavior Cloning for Quake with Frozen Visual Features ​

Author: Dzmitry Malyshau
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.22739v2 Announce Type: replace-cross Abstract: We study how far a deliberately simple behavioral-cloning policy can progress in a visually rich first-person game before adding reinforcement learning or explicit memory. Cortex is a compact Quake policy with 10.98 million trainable paramete...

📖 Read original article


288. WCM: World-Cognition Model for Generalizable Human-Robot Interaction ​

Author: Yuzhen Chen, KC Zhou
Published: 8/5/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.HC, cs.LG

arXiv:2607.22999v2 Announce Type: replace-cross Abstract: Language agents can now interact fluently with users in software, but robots still struggle to bring comparable interaction to physical tasks. Current robot-control paradigms, including vision-language-action policies and world-model-based pl...

📖 Read original article


289. CAPT: A Multi-task Continuous Autoregressive Transformer enabling Cross-dataset and Cross-species Transfer for Calcium Population Dynamics ​

Author: Xinhong Xu, Yimeng Zhang, Yuanlong Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.23258v2 Announce Type: replace-cross Abstract: Large-scale calcium imaging has created an opportunity to build foundation-style models for neural population dynamics, but a central question remains unresolved: \textbf{whether a model pretrained on one collection of recordings can generali...

📖 Read original article


290. ForgettingOT: Certified Speculative Batching from Sinkhorn's Projective Forgetting ​

Author: Xinyang Wen
Published: 8/5/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2607.24741v2 Announce Type: replace-cross Abstract: Positive two-marginal entropic optimal transport is solved by a nonlinear, positive, order-preserving, homogeneous Sinkhorn map. After quotienting the dual scaling gauge, we show that the active eigenmode of the fixed-point Jacobian $J_t^\sta...

📖 Read original article


291. Bumblebee: Interleaved Mixed-Layer Building Blocks for Large-Scale Recommendation Systems ​

Author: David Bauer, Cancan Zhang, Wenshun Liu, Xiaoyi Zhang, Weijia Liu, Wanli Ma, Yue Weng, Wei Li, Rui Li, Jing Qian, Huayu Li, Xiaoyi Liu, Linhong Zhu, Jerry Fu
Published: 8/5/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.24804v2 Announce Type: replace-cross Abstract: Recommendation systems have undergone significant transformations in the past years. The transition from traditional feature interaction modules to generative next-action prediction has pushed the boundaries of personalized content. Developme...

📖 Read original article


292. Learning the Word Problem: Geodesic Lengths and Cryptographic Applications ​

Author: Elisabeth Fink
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, math.GR

arXiv:2607.26241v2 Announce Type: replace-cross Abstract: The Word Problem has been a subject of intensive mathematical study for over a century, initially driving advances in combinatorial group theory and more recently emerging as a foundational hardness assumption in post-quantum cryptography (PQ...

📖 Read original article


293. Metis: Memory Foundation Model ​

Author: Zeyu Zhang, Ziliang Guo, Yihang Sun, Xichong Zhang, Xixuan Hao, Zehao Lin, Yang Zhang, Xiaoyan Zhao, Tong Shen, Bo Tang, Zhi-Qin John Xu, Junchi Yan, Haofen Wang, Xu Chen, Feiyu Xiong, Zhiyu Li, Tat-Seng Chua
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.26760v2 Announce Type: replace-cross Abstract: Recent advances in AI agents have increasingly internalized native capabilities into their underlying foundation models, giving rise to multimodal foundation models and large reasoning models. However, agent memory is still primarily implemen...

📖 Read original article


294. Feature Bagging Provides Stability ​

Author: Yuheng Ma, Qiang Sun
Published: 8/5/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2607.26964v2 Announce Type: replace-cross Abstract: We study feature bagging through the lens of algorithmic stability. Feature bagging is an ensemble strategy that aggregates base learners trained on randomly subsampled feature subsets, possibly in a data-dependent manner. We introduce featur...

📖 Read original article


295. AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents ​

Author: Ruoyu Wang, Heng Zhao, Renjie Wu, Mengnan Zhao, Zhixuan Chu, Wanyu Lin, Tianhang Zheng
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG

arXiv:2607.26998v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents automate penetration testing through an observation-action loop, selecting actions based on observations returned by tools. This dependence allows defenders to inject deceptive observations that can mislead t...

📖 Read original article


296. Investigating reservoir computing for branch prediction in pipelined processors using emerging CMOS memristor devices ​

Author: Harvey Samuel George Johnson, Sendy Phang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AR, cs.CE, cs.ET, cs.LG, physics.app-ph

arXiv:2607.27140v2 Announce Type: replace-cross Abstract: This project aimed to develop a novel reservoir compute (RC) implementation framework targeting high-speed operation and integration with CMOS digital logic. With the target workload of branch prediction (BP) for multistage pipelined central ...

📖 Read original article


297. Automated ECG Interval Measurement and Wave Delineation Using Fast Fourier Convolution ResNet ​

Author: Farhan Adam Mukadam, Harshit Mishra, Nachiket Makwana, Pradyot Tiwari, Subramani Kandasamy, KVS Hari
Published: 8/5/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG

arXiv:2608.00058v2 Announce Type: replace-cross Abstract: Accurate measurement of ECG intervals, including PR, QRS duration, and QT/QTc, is central to cardiac diagnosis, yet the published ECG delineation literature evaluates performance almost exclusively as fiducial-point timing errors on small cur...

📖 Read original article


298. LLM-OSDA: An Optimal-Stopping Dynamic Auction for Native Advertising in Multi-Turn LLM Conversations ​

Author: Yan Fang, Jialin Chen, Chun Gan, Hang Yu, Mingjun Nie, Yeyu Zhang, Fengxiang He, Ching Law
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.GT, cs.LG

arXiv:2608.00123v2 Announce Type: replace-cross Abstract: LLM-native advertising embeds sponsored content directly into model-generated responses, shifting the unit of sale from a fixed slot to a moment within an evolving conversation. Existing LLM ad-auction mechanisms primarily operate within a si...

📖 Read original article


299. A Constitution-Grid Instrument for Data-Efficient RL Alignment (C-Guard) ​

Author: Xianling Zhang
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.00180v2 Announce Type: replace-cross Abstract: Conflicting objectives are general in RL alignment, and training on them data-efficiently is hard. Training a safety guard with RL means optimizing two objectives that conflict: catch real harm, and do not refuse benign prompts. Our finding i...

📖 Read original article


300. Pruned BPE: Post-training Visibility Pruning and Token Reallocation for Byte Pair Encoding ​

Author: Kenny Shao
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.00837v2 Announce Type: replace-cross Abstract: Byte Pair Encoding (BPE) is widely used for subword tokenization, but standard BPE exposes every learned merge token to the downstream model, including tokens that mainly serve as intermediate construction units and rarely appear in the final...

📖 Read original article


301. LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing ​

Author: Wen Zan, Jiaqi Zhang, Jianchao Tan, Hong Liu, Cunguang Wang, Xiang Li, Duyue Ma, Guanyu Wu, Yifan Lu, Fengcun Li, Yerui Sun, Peng Pei, Yuchen Xie, Xunliang Cai
Published: 8/5/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.DC, cs.LG

arXiv:2608.01662v2 Announce Type: replace-cross Abstract: DeepSeek Sparse Attention (DSA) enables efficient long-context modeling through its Lightning Indexer. However, practical deployment remains constrained by the indexer's expensive $O(L^2)$ scoring overhead and the hardware-inefficient, discon...

📖 Read original article


302. Self-Improving Large Language Models via Progressive Experience Evolution ​

Author: Shijie Ren, Xiting Wang, Meng Li, Yujie Guo, Yunhang Yao, Ziheng Peng, Xunlong Wang, Yuetan Chen, Haoyang Zhou, Yunlong Liang, Fandong Meng
Published: 8/5/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.02139v2 Announce Type: replace-cross Abstract: Large language models (LLMs) capable of self-improvement require not only effective policy optimization, but also a principled mechanism for transforming transient interaction experience into persistent model capabilities. Existing self-impro...

📖 Read original article