Skip to content

arXiv cs.LG - 2026-08-12 ​

295 items collected.


1. Transformer Geometry Observatory TGO-IV: Developmental Topology Observatory ​

Author: Kaustubh Kapil, Kishor P. Upla
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.09997v1 Announce Type: new Abstract: Transformers have had a profound impact on the world of language processing and computer vision. As efforts to answer the million-dollar question of ``How does a Transformer learn?" have been increasing, existing interpretability studies primarily anal...

📖 Read original article


2. Uncertainty-Aware Ensemble Deep Randomized Neural Networks for Classification ​

Author: M. Sajid, A. Quadir, A. Rahaman, P. N. Suganthan, M. Tanveer
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10007v1 Announce Type: new Abstract: The current state-of-the-art (SOTA) deep randomized neural networks, such as deep Random Vector Functional Link (dRVFL) and ensemble deep RVFL (edRVFL), treat all training samples uniformly, which limits their robustness and effectiveness when applied ...

📖 Read original article


3. CurveFP: Co-Designing Numerical Representation and Product Arithmetic for Language Models ​

Author: Ye Qiao
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10010v2 Announce Type: new Abstract: Low-precision formats usually optimize scalar fidelity while inheriting conventional product arithmetic. We introduce CurveFP, a block-scaled family that distributes magnitudes across interleaved logarithmic curves. Uniform curve indices make every non...

📖 Read original article


4. Sheaf-Based Federated Representation Learning ​

Author: Gabriele D'Acunto, Enrico Grimaldi, Valeria Avino, Mario Edoardo Pandolfo, Leonardo Di Nino, Sergio Barbarossa, Paolo Di Lorenzo
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA, eess.SP

arXiv:2608.10016v1 Announce Type: new Abstract: Heterogeneous federated systems require agents to learn and exchange informative representations despite differences in data distributions, sensing modalities, model architectures, latent dimensionalities, and local learning objectives. To address this...

📖 Read original article


5. DOCSCHISEL: Adaptive Tool Documentation Optimization Framework for LLM Agents ​

Author: You Lu, Kun Zhang, Bihuan Chen, Xin Peng
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10037v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on external tools to accomplish complex real-world tasks, making tool documentation a critical grounding resource for LLM agents. Existing studies mainly focus on improving the tool-use capabilities of LLM...

📖 Read original article


6. FlowScout: From Execution Feedback to Reliable Tool-Using Agent Workflows ​

Author: Shuo Hao, You Lu, Bihuan Chen, Xin Peng
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10039v1 Announce Type: new Abstract: Agentic workflows have become an important abstraction for building reliable LLM-based automation systems by organizing large language models (LLMs), tools, and control logic into explicit execution structures. However, constructing high-quality agenti...

📖 Read original article


7. UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs ​

Author: Xuexiong Yin, Zechuan Chen, Yongsen Zheng, Yuxiang Zhang, Jingyuan Yang, Bin Wang, Yubin Wang, Keze Wang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10042v1 Announce Type: new Abstract: Tool-use LLMs are increasingly asked to act on users' behalf, but existing benchmarks usually focus on profile recall, style imitation, generic tool use, or response-level personalization. We introduce UserToolBench , a benchmark for personalized decis...

📖 Read original article


8. Finding the Signal in the Spam: Jointly Learning Rewards and Worker Reliability from Pairwise Comparisons ​

Author: Kaustubh Shivshankar Shejole, Tanish Agarwal, Arpit Agarwal, Avishek Ghosh
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10045v1 Announce Type: new Abstract: The problem of learning from pairwise comparisons has been widely studied across many domains such as recommendation systems, social choice, and more recently, fine-tuning large language models. In this problem, the goal is to learn item rewards based ...

📖 Read original article


9. Detecting Soft Skills in ML Engineering Roles CVs ​

Author: Aidin Azamnouri, Nouran Ayad, Justus Bogner, Stefan Wagner
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2608.10046v1 Announce Type: new Abstract: Soft skills shape collaboration among ML engineers, data scientists, and software engineers building ML-enabled systems, yet what we know about them comes almost entirely from the demand side. Job advertisements, surveys, and hiring manager interviews ...

📖 Read original article


10. Physics-Informed Machine Learning in Prognostics and Health Management: A Systematic Literature Review ​

Author: Christopher Braun, Julian Raible, Marco F. Huber
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10047v1 Announce Type: new Abstract: In modern industry, keeping complex systems reliable, safe, and efficient hinges on Prognostics and Health Management (PHM). Machine Learning (ML) has largely driven advancements in diagnostics and prognostics, yet purely data-driven models face inhere...

📖 Read original article


11. Observational Policy Ranking for SMB Financial Guidance from Multi-Action Accounting Logs ​

Author: Shrutendra Harsola, Vignesh Subrahmaniam, Vikas Raturi, Kamalika Das, Xiang Gao, Kratika Gupta, Ruocheng Guo, Padmaja Jonnalagedda, Ananya Pramod, Sricharan Kumar
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10050v1 Announce Type: new Abstract: Small and medium-sized businesses need timely financial guidance, yet historical accounting logs record self-selected and often co-occurring business changes rather than randomized recommendations. We formulate this setting as observational policy rank...

📖 Read original article


12. ChronoSSM: Training for Temporally Aware Representations in Autoregressive State Space Models ​

Author: Adrien Schoen, Nachiketa Ratnakar Patil, Arjun Bhagoji, Francesco Bronzino
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.NI

arXiv:2608.10120v1 Announce Type: new Abstract: Modern sequence models, from Transformers to State Space Models, have enabled powerful generative modeling across diverse domains, yet they are typically trained to predict what happens while treating when it happens as a secondary concern. In data-min...

📖 Read original article


13. Procedural Fairness Failures in RLHF from Preference Averaging ​

Author: M P V S Gopinadh, Karthik Kamuju, Kummari Avinash, John Joshua, Srinivasa Raju Rudraraju
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.10126v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) aggregates heterogeneous preferences into a single reward model, assuming preference homogeneity. When preferences are heterogeneous, this aggregation induces a procedural fairness failure where majorit...

📖 Read original article


14. SeFoRA: Sketch-Aggregated Federated Low-Rank Adaptation with Heterogeneous Client Ranks ​

Author: Yue Xia, Tayyebeh Jahani-Nezhad, Mayank Bakshi, Rawad Bitar
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2608.10144v1 Announce Type: new Abstract: We consider federated parameter efficient fine-tuning of large neural networks with low-rank adaptation (LoRA,~Hu et al.\ 2022). Combining LoRA with federated PEFT introduces challenges absent from either setting alone: clients may use different LoRA r...

📖 Read original article


15. The Evaluation Protocol Determines the Result: An Independent Reproduction of LeWorldModel on TwoRoom ​

Author: Joyjeet Singh
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10145v1 Announce Type: new Abstract: LeWorldModel trains a latent world model with a prediction loss and a single anti-collapse regulariser, and reports approximately 87% of goals reached on TwoRoom, its simplest diagnostic environment. We reproduce that result by independent reimplementa...

📖 Read original article


16. REATS: LLM Reasoning-based Ensemble Learning for Adaptive Time Series Forecasting ​

Author: Xu Zhang, Chang Xu, Hui Sun, Nan Ma, Zijian Zhang, Peng Wang, Wei Wang, Li Zhao
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10149v1 Announce Type: new Abstract: Due to the diversity of real-world time series, no single forecasting model consistently dominates across all samples. Ensemble learning addresses this by combining complementary model strengths, yet existing methods rely on fixed rules or black-box mo...

📖 Read original article


17. Intrinsic Structure: Spectral Identifiability for Mechanistic Interpretability ​

Author: Ashim Dhor, Pin-Yu Chen
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10172v1 Announce Type: new Abstract: Mechanistic interpretability explains models by identifying circuits inside them, but has no way to tell whether a circuit is a property of the model or an artifact of the method that found it. Sparse autoencoders illustrate the problem: different seed...

📖 Read original article


18. From Prediction to Incrementality: Causal Optimization for Large-Scale Targeting and Recommendation ​

Author: Changshuai Wei, John Bencina, Phuc Nguyen, Andre Assuncao Silva T Ribeiro, Benjamin Zelditch
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10182v1 Announce Type: new Abstract: Large-scale targeting and recommendation systems are typically built around predictive scores fed into heuristic or local allocation. When the business goal is incremental impact, as in marketing campaigns, incentives, and notifications, this paradigm ...

📖 Read original article


19. ELMER: Evolutionary Language Model that Explores and Refines ​

Author: Matthew Siper, Ahmed Khalifa, Julian Togelius
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10196v1 Announce Type: new Abstract: Program evolution can measure whether a mutation helped, but it rarely controls how far the mutation moves in behavior space. Syntactic edit size is an unreliable proxy: a small code change can alter nearly every action, while a larger rewrite can pres...

📖 Read original article


20. Boundary-Seeking Policy Gradient for Safe Reinforcement Learning ​

Author: Chenhua Fan, Jiahui Zhu, Yuhang Zhang, Honghao Wei
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10204v1 Announce Type: new Abstract: Safe reinforcement learning maximizes reward subject to safety constraints. For Constrained Markov Decision Processes, the linear-programming view over occupancy measures implies that whenever the constraint is active at optimality, the optimal policy ...

📖 Read original article


21. A matched-integrator evaluation of Hamiltonian neural networks on pendulum and Kepler dynamics ​

Author: Lenick Kemunto Nyabuto, Yae Ulrich Gaba, Birahim Tewe
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10235v1 Announce Type: new Abstract: Hamiltonian Neural Networks (HNNs) parameterize conservative dynamics through a learned scalar Hamiltonian, providing an architectural prior that is absent from generic vector-field neural networks. We evaluate this prior under a controlled protocol in...

📖 Read original article


22. STCAD: Scalable Trajectory Clustering and Anomaly Detection on Terabyte-Scale AIS Data ​

Author: Bertram Hage, Alexander Schi{\o}tz, Felix Thomsen, Christian Rand, Peder Heiselberg
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10249v1 Announce Type: new Abstract: We present a scalable framework for unsupervised clustering of maritime trajectories derived from terabyte-scale Automatic Identification System (AIS) archives. Variable-length trajectories are encoded with a custom BERT-based model trained via masked ...

📖 Read original article


23. CRHT: A Continuous Regression Hybrid Transformer for Vessel Trajectory Prediction with Online Cluster Sampling ​

Author: Alexander Schi{\o}tz, Bertram Hage, Christian Rand, Felix Thomsen, Peder Heiselberg
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10256v1 Announce Type: new Abstract: Accurate vessel trajectory prediction is critical for maritime safety and anomaly detection, yet existing models often struggle with geographic bias and navigational realism. We propose the Continuous Regression Hybrid Transformer (CRHT), a deep learni...

📖 Read original article


24. Toward Human Rights Benchmarking for LLMs: A Pilot Methodology ​

Author: Savannah Thais, Wm. Matthew Kennedy, Abhigyan Acherjee, Matilda Wysocki, Malcolm Langford, Caitlin Kraft Buchman
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY

arXiv:2608.10268v1 Announce Type: new Abstract: Large language models (LLMs) increasingly mediate legal determinations over what human rights are realized, and how. Yet, no evaluation benchmark exists to assess whether they can reason correctly about human rights law. To this end, we report our effo...

📖 Read original article


25. Power law graph attention: exact generalization of scaled dot-product attention, empirical collapse at inference ​

Author: Burc Gokden
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.10288v1 Announce Type: new Abstract: The Large Language Model from Power Law Decoder Representations (PLDR-LLM) and its attention, Power Law Graph Attention (PLGA), replace the fixed bilinear form of scaled dot-product attention (SDPA) with a learned, input-generated bilinear operator $G_...

📖 Read original article


26. MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at Scale ​

Author: Yuhang Yao, Zeyu Wang, Wanyi Chen, Tongyun Yang, Yuhang Han, Jie Xiao, Chengke Bao, Tianyi Zhao, Lynn Ai, Eric Yang, Tianyu Shi
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10333v1 Announce Type: new Abstract: LLM agents execute heterogeneous sequences of model calls within a single task: some invocations require careful reasoning, while others are structured steps such as formatting or tool-argument construction. Prior routing methods exploit this asymmetry...

📖 Read original article


Author: Karl Pierce, Yuehaw Khoo, Haizhao Yang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10351v1 Announce Type: new Abstract: In this work we present a method to accelerate the optimization of learning high dimensional functions using deep neural network (DNN). This optimization procedure introduces contextual features into the first layer of a DNN. The parameters of DNN are ...

📖 Read original article


28. Efficient Reinforcement Learning for Long-Horizon Tool-Use Agentic Tasks ​

Author: Zelei Cheng, Amritansh Mishra, Sambit Sahu, William Campbell
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10357v1 Announce Type: new Abstract: Long-horizon tool-using agents must reason over user goals, domain policies, tool calls, simulator state, and delayed verifiable rewards. Reinforcement learning (RL) is a natural fit for this setting, but multi-turn on-policy rollouts create long conte...

📖 Read original article


29. Invertible Logits Transformation for Accuracy-Preserving Post-Hoc Uncertainty Calibration ​

Author: Lening Zhao, Qipeng Zhan, Li Shen
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10372v1 Announce Type: new Abstract: Post-hoc calibration aligns a classifier's predicted confidences with its empirical accuracy without retraining. An ideal calibrator should correct nonlinear miscalibration, scale gracefully to large label spaces, and preserve the original predictions;...

📖 Read original article


30. Fisher8: Stabilizing Neural Heteroscedastic Regression via Output-Layer Fisher Geometry ​

Author: Sumedh Vemuganti, Nickvash Kani
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10374v1 Announce Type: new Abstract: Training neural networks to jointly predict mean and uncertainty estimates from noisy observations can be unstable, prompting a series of independent stabilization efforts. We argue that these interventions highlight a common underlying issue where gra...

📖 Read original article


31. Generator-Guided Inverse Sampling for L\'evy-Driven Generative Models ​

Author: Tianfu Qi, Jun Wang, Jun Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10384v1 Announce Type: new Abstract: This paper studies inverse sampling for L'evy-driven generative models from the perspective of Markov generators. Unlike conventional diffusion models, L'evy-driven dynamics involve infinite jump activities, which makes their reverse process nonlocal...

📖 Read original article


32. Dreamer-SAC: Off-Policy Learning in Latent World Models for Sample-Efficient Autonomous Driving ​

Author: Jiazhuo Li, Linjiang Cao, Qi Liu, Xi Xiong
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.10386v1 Announce Type: new Abstract: Sample-efficient reinforcement learning for autonomous driving is often limited by the trade-off between data efficiency and model bias. While world models reduce the reliance on costly environment interactions, policy optimization over learned dynamic...

📖 Read original article


33. Share First, Route What Remains: A Unified Framework for Token-Adaptive MoE Computation ​

Author: Gongli Zhang, Zhulin Liu, C. L. Philip Chen
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CV

arXiv:2608.10392v1 Announce Type: new Abstract: Mixture-of-experts (MoE) models have recently moved beyond routing a fixed number of complete experts. Shared-expert designs preserve reusable knowledge, fine-grained methods vary computation within experts, and dynamic routers adapt the number of acti...

📖 Read original article


34. ELVAE: Evidential Learning-Based Variational Autoencoder for Uncertainty-Aware Generation ​

Author: Ge Wang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10398v1 Announce Type: new Abstract: Variational autoencoders generate samples from probabilistic latent representations but do not distinguish uncertainty about the latent location from variability around it. We formulate ELVAE, an evidential learning-based VAE in which each latent coord...

📖 Read original article


35. Do Judges Behave Like Algorithms? ​

Author: Riya Manchanda, Eric Chen, Chloe Zhu, Cynthia Rudin, Brandon Garrett, Songman Kang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10400v2 Announce Type: new Abstract: What if judges already behave like algorithms? As artificial intelligence and algorithms are deployed in many settings, including the judicial system, many have debated whether judges should be allowed to rely on them. Instead, we ask whether judges fo...

📖 Read original article


36. TideRL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling ​

Author: Yanyu Ren, Xizheng Wang, Xiao Liu, Bowen Lv, Hanchen Zhang, Shudan Zhang, Hanyu Lai, Shuai Wang, Li Chen, Dan Li, Jie Tang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2608.10402v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models is moving toward multi-turn agentic workloads, where rollout tasks repeatedly pause for external environments, resume with growing contexts, and finish at highly variable times. In this setting, RL ...

📖 Read original article


37. Actionable Hallucination Detection: Translating Latent Uncertainty into Agentic Critique ​

Author: Sanidhya Vijayvargiya, Rahul Lokesh
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10430v1 Announce Type: new Abstract: Large Language Models (LLMs) deployed as AI agents frequently exhibit user specification-grounding failures, executing hallucinated, undesired actions to force a resolution rather than expressing uncertainty. Existing detection methods fail to provide ...

📖 Read original article


38. Do Time-Series Forecasters Use the Right History: Recoverability, Recovery, and Functional Use of Temporal Delays ​

Author: Qipeng Qian, Yuntao Qian
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10433v1 Announce Type: new Abstract: Forecast accuracy does not tell us which past inputs produced a prediction. We separate three questions for time-series models with known delay structure: can the true delay be recovered from the observed data, does the model report it, and does the fo...

📖 Read original article


39. Detecting an Effect Is Not Learning to Act on It: A Reward-SNR Floor for LLM Acquisition Agents ​

Author: Ying Yuan
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IR

arXiv:2608.10441v1 Announce Type: new Abstract: Many pipelines can pay a per-example cost to acquire an auxiliary, model-derived observation -- an LLM's structured reasoning, a slow oracle, an expensive measurement -- and then must decide when the acquired signal is worth using. Our thesis is a dist...

📖 Read original article


40. A Joint-Distribution Route to Fair Representations with Continuous Sensitive Attributes ​

Author: Yijin Ni, Xiaoming Huo
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, stat.AP

arXiv:2608.10470v1 Announce Type: new Abstract: Fair representation learning with a continuous sensitive attribute $S$ requires a representation $Z$ that is statistically independent of $S$. Existing criteria, including generalized demographic parity, the expectation of integral probability metrics ...

📖 Read original article


41. Critic-Free Pretraining for Efficient Online Reinforcement Learning Fine-Tuning ​

Author: Daoyi Li, Yixian Zhang, Chao Yu, Wenbo Ding, Yu Wang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10473v1 Announce Type: new Abstract: Offline-to-online (O2O) reinforcement learning aims to leverage policies pretrained on static datasets while improving them through online interaction. However, directly reusing an offline-trained critic can hinder online fine-tuning: as the policy and...

📖 Read original article


42. Exploration-Driven Personalized Federated Reinforcement Learning via Intrinsic Motivation ​

Author: Md Rafid Islam, Rafsan Jany, Zahid Hasan, Ratun Rahman
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10499v1 Announce Type: new Abstract: Personalized Federated Reinforcement Learning (PFRL) takes a decentralized approach to storing and accessing information based on past experiences while keeping each client's data private during the learning of each client's policy. Many current method...

📖 Read original article


43. Coordinating the Unknown Lipschitz Constant in Multiplayer Bandits ​

Author: Ricardo Parada, Chenzhang Zhao, William Chang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10526v1 Announce Type: new Abstract: Motivated by decentralized applications, we study cooperative multi-agent bandits in continuous (Lipschitz) action spaces when the Lipschitz constant is unknown. We consider three information structures: (A)~unobserved actions with common rewards, (B)~...

📖 Read original article


44. Robust Multi-Agent Bandits with Heavy-Tailed Rewards and Information Asymmetry ​

Author: Daphne Feng, Ricardo Parada, Lily Jiang, Sophia Yi, William Chang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10529v1 Announce Type: new Abstract: The multi-armed bandit problem is a central framework in sequential decision-making, extensively studied under sub-Gaussian reward assumptions. However, real-world applications often involve heavy-tailed reward distributions and decentralized, informat...

📖 Read original article


45. Retrieval-Corrected Conformal Prediction for Time Series ​

Author: Sangjin Jin, Kangmin Kim, Junhyeong Lee, Yongjae Lee
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10553v1 Announce Type: new Abstract: Conformal prediction (CP) provides distribution-free prediction intervals for fixed forecasters, but its standard calibration procedure is often inefficient for time series data, where forecast errors are temporally dependent and change across time and...

📖 Read original article


46. MARCO: Click-Intent Decomposition for Calibrated Ads Conversion Prediction ​

Author: Shiwen Shen, Xiru Huang, Liang Luo, Jianbo Sun, He Lyu, Zihang Fu, Ivonne Xu, Zhizhuo Li, Zhengyu Zhang, Pei-Ju Sung, Yunmiao Wang, Zixuan Wang, Zhengli Zhao, Qiang Jin, Mike Jermann, Mingda Li, Yang Xiao, Bhavana Challa, Brooke Bian, Yang Li, Ashish Chamoli, Bibek Bhusal, Danning Di, Yuan Jin, Meet Raval, Zhiwen Chen, Boyao Sun, Shuguang Wang, Yunlong He, Yantao Yao, Sagar Chordia, Wenlin Chen, Santanu Kolay, Qin Huang, Ellie Wen
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10562v1 Announce Type: new Abstract: Not all clicks are equal. Industrial ads ranking decouples conversion probability into click-through rate (CTR) and post-click conversion rate (CVR), yet treats every click as the same event. In reality, users provide a free, self-generated signal of i...

📖 Read original article


47. BREAD: Baseline-Referenced Explanations for Anomaly Diagnosis ​

Author: Jiaqi Qiu, Rob Goedhart, Jannis Kurtz, Inez M. Zwetsloot
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10587v1 Announce Type: new Abstract: Artificial Intelligence (AI)-based prospective anomaly detection methods are increasingly deployed in high-dimensional and nonlinear settings. Among these approaches, AI-based statistical process monitoring (SPM) is widely used, providing a structured ...

📖 Read original article


48. $\beta$-VAEs as Effective Theories: Tolerance-Dependent Dimension ​

Author: Johannes Hirn
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10599v1 Announce Type: new Abstract: In a $\beta$-VAE, increasing the regularization strength acts as a spectral cutoff by collapsing low-utility latent coordinates. In the linear Gaussian VAE, the collapse order matches the ranking of reconstruction utilities exactly, because both are se...

📖 Read original article


49. Compute-Optimal Is Not Cluster-Optimal: Systems-Aware Scaling for Sparse Mixture-of-Experts ​

Author: Soumajyoti Sarkar, Yuxin Tang, Sheng Zha
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10605v1 Announce Type: new Abstract: In large-scale pretraining, the algorithm, architecture, and systems decisions are conventionally made in disconnected stages. A scaling law stage selects an architecture and training recipe, optimizing loss under compute constraints, and a separate sy...

📖 Read original article


50. Pair-Centric Graph Rewiring for Over-Squashing via Optimal Transport-Guided Communication Alignment ​

Author: Yan Wang, Chuan-Xian Ren
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10619v1 Announce Type: new Abstract: Message-passing neural networks (MPNNs) often struggle when task-relevant information is distributed across distant regions of a graph, since local propagation must compress remote signals through limited structural interfaces. Graph rewiring provides ...

📖 Read original article


51. ProbGuard: Calibrated Safety Risk Estimation from LLM Output Distributions ​

Author: Xinzhe Huang, Biwu Yao, Kedong Xiu, Mengnan Zhao, Di Wang, Puning Zhao, Tianhang Zheng
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10621v1 Announce Type: new Abstract: Recent research on Large Language Model (LLM) safety has widely adopted guardrails to identify unsafe LLM outputs. Existing guardrails typically formulate safety assessment as a deterministic classification task, mapping a discrete token sequence to a ...

📖 Read original article


52. IADD-TR: Intervention-Aware Dynamics Decoupling with Targeted Regularization for Model-Based Reinforcement Learning ​

Author: Zefeng Liang, Jie Qiao, Ruichu Cai, Weilin Chen, Zhifeng Hao
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10634v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL), which learns environment dynamics to generate synthetic experience, is a promising approach to sample-efficient decision making. Numerous methods have been developed to improve dynamics prediction and policy o...

📖 Read original article


53. Optimize Cheap, Deploy Strong: Cost-Aware Cross-Tier Transfer for Evolutionary Optimization ​

Author: Tal Oved, Roi Pony, Oshri Naparstek, Udi barzelay
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.NE

arXiv:2608.10694v2 Announce Type: new Abstract: Evolutionary optimization of LLM prompts and agentic programs (e.g., GEPA) is dominated by fitness evaluation: scoring each candidate runs an answering LLM over a validation set, so the evaluator's price tier dictates total search cost. We restructure ...

📖 Read original article


54. ProTAGAD: A Foundation Model for TAG Anomaly Detection with Decoupled Topological and Textual Prototypes ​

Author: Ziyan Wang, Liwen Wu, Cheng Xie, Song Gao, Zhenli He, Xin Jin
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10699v1 Announce Type: new Abstract: Text-Attributed Graphs (TAGs), endowed with abundant textual content along with topological structures, have emerged as a versatile backbone for real-world anomaly detection spanning large language model security, social network moderation, and cyber t...

📖 Read original article


55. Your LLM, Your Style: Behavioral Mode Axes for LLM Behavioral Control ​

Author: Haoze Liu, Run Liu, Haiying Xu, Jiahui Han, Siyuan Fang, Siyu Yan, Huiqi Deng, Guanchu Wang, Na Zou
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.HC

arXiv:2608.10703v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act in interactive settings where their behavioral styles affect user experience, safety, and downstream decision making. Existing LLM personality studies largely rely on self-report questionnaires administered...

📖 Read original article


56. SQuaT: Self-Supervised Knowledge Distillation via Student-Aware Quantized Teacher Features ​

Author: HyeonJun Lee, Hyeonsik Jo, Jinwoo Chung, Jangho Kim
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10709v1 Announce Type: new Abstract: Quantization-Aware Training (QAT) enables the deployment of quantized models with minimal accuracy degradation. However, in practical scenarios, training labels are often unavailable due to privacy, copyright, or cost constraints. Knowledge Distillatio...

📖 Read original article


57. Long-Time Trajectory Approximation via SA-NODEs: Model Predictive and Floquet Strategies ​

Author: Ziqian Li, Nikolaos M. Matzakos
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2608.10738v1 Announce Type: new Abstract: We study the approximation of dynamical systems by semi-autonomous neural ordinary differential equations (SA-NODEs) over long time horizons. For a single network trained on the whole horizon, the available error bound deteriorates double exponentially...

📖 Read original article


58. Path Integral Value Matching for Linear Quadratic Stochastic Optimal Control ​

Author: Bangyan Liao, Chenglei Yu, Yuchen Yang, Chuanrui Wang, Zhisheng Song, Peidong Liu, Tailin Wu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2608.10777v1 Announce Type: new Abstract: Linear Quadratic Stochastic Optimal Control (LQ-SOC) establishes a fundamental framework for steering noisy dynamical systems and has recently gained renewed interest in the machine learning community. However, current state-of-the-art policy-based met...

📖 Read original article


59. MoE Proxy Models for Low-Cost Failure Reproduction and Diagnosis in LLM RL Post-Training ​

Author: Yikai Wang, Chuansai Zhou, Yuhang Zhou, Weiqiang Wu, Cong Wu, Yue Deng, Ben Feng, Mingming Zhu, Beirong Zhou, Zhibin Wang, Sheng Zhong, Chen Tian, Wangze Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10823v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training of large language models (LLMs) is computationally intensive and involves complex system pipelines with substantial debugging overhead. In practice, factors such as framework adaptation, numerical precision, an...

📖 Read original article


60. TACTICL: Task-Aware Compression of Tabular ICL Models ​

Author: Mykhailo Koshil, Matthias Feurer, Katharina Eggensperger
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.10837v1 Announce Type: new Abstract: The strong performance of foundation models for tabular tasks comes at substantial inference costs. Distilling models into task-specific architectures reduces model size and computational demands but also sacrifices in-context adaptability. Here we int...

📖 Read original article


61. Diffract: Spectral View of LLM Domain Adaptation ​

Author: Nikita Borodin, Maria Krylova, Artem Zabolotnyi, Dmitry Aspisov, Egor Shikov, Nikita Tyuplyaev, Oleg Travkin, Roman Alferov, Dmitry Vinichenko
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10850v1 Announce Type: new Abstract: We study continual pre-training (CPT) as a mechanism for adapting general-purpose large language models to specialized domains: mathematics, instruction, code, and natural text. Using singular value decomposition of weight matrices, we find that CPT le...

📖 Read original article


62. FiGuRO: Intrinsic Dimension Estimation for Multi-Modal Data ​

Author: Viktoria Schuster, Sana Tonekaboni, Caroline Uhler
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10857v1 Announce Type: new Abstract: Determining the complexity, or Intrinsic Dimension (ID), of data is fundamental to efficient and interpretable representation learning. This is particularly challenging in multi-modal settings when trying to learn disentangled representations for share...

📖 Read original article


63. Can Bayesian Optimization Efficiently Find a Strong Single Expert in Neural Thickets? ​

Author: Nigel Bastian Cendra, Abdelhamid Ezzerg, Fernando Julio Cendra, Jeremias Knoblauch, Jakob Zeitler
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10867v1 Announce Type: new Abstract: Gradient-free post-training has emerged as a compelling alternative to gradient-based optimization for large language models (LLMs), but existing approaches remain costly. We ask whether structured search can identify a strong single expert under a mod...

📖 Read original article


64. Optimistic Rates for Multiclass PAC Learning ​

Author: Xiaoyu Li, Andi Han, Jiaojiao Jiang, Junbin Gao
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.10869v1 Announce Type: new Abstract: Worst-case multiclass bounds do not become smaller when the best classifier is already nearly correct: what is missing is an optimistic rate, a guarantee whose fluctuation scales with the oracle risk itself. For a class of Natarajan dimension $d_N$ and...

📖 Read original article


65. Benchmarking Time Series Generation Methods for Privacy-Preserving Forecasting ​

Author: Luis Amorim, Vitor Cerqueira, Moises Santos, Paulo J. Azevedo, Carlos Soares
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10891v1 Announce Type: new Abstract: Time series forecasting in privacy-sensitive domains often requires training models on released data rather than original observations. Synthetic time series generation has been developed primarily for data augmentation, where generated series suppleme...

📖 Read original article


66. Partially Observable Learning for Multi-Platform Dispatch Optimization ​

Author: Fengming Yao, Man Luo
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10897v1 Announce Type: new Abstract: Instant delivery platforms have become a critical component of urban logistics, increasingly relying on crowdsourced couriers to fulfill highly dynamic orders. In real-world systems, couriers are not exclusive to a single platform and may concurrently ...

📖 Read original article


67. ReOrder-OPD:Reliability-Aware Prompt Ordering for On-Policy Distillation ​

Author: Ximo Zhu, Ruiqi Liu, Rong Wang, Ping Wu, Xiang Zheng, Wenzhuo Xu, Xubin Yao, Zhiyuan Yan, Bo Li, Jun Gao, Xiaolei Lv
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10905v1 Announce Type: new Abstract: On-policy distillation (OPD) applies token-level teacher supervision to student-generated trajectories, but this supervision is not always reliable. Existing methods use local confidence or teacher-student agreement to weight, filter, or truncate the s...

📖 Read original article


68. Physics-informed Diffusion Generative Model for Time-Series Data Synthesis in Dynamic Systems ​

Author: Haiteng Wang, Yunfei Zhu, Tao Wang, Yikang Li, Jiabao Dong, Xiaoge Zhang, Lei Ren
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10941v1 Announce Type: new Abstract: Industrial time-series signals, such as turbine temperature and rotational speed in aero-engines, are essential for monitoring the health and operational status of complex dynamical systems. However, collecting such data is often limited by harsh envir...

📖 Read original article


69. GARLIC: Graph Attention-based Relational Learning of Multivariate Time Series in Intensive Care ​

Author: Ruirui Wang, Yanke Li, Manuel G"unther, Diego Paez-Granados
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10969v1 Announce Type: new Abstract: Healthcare data, such as Intensive Care Unit (ICU) records, comprise heterogeneous multivariate time series sampled at irregular intervals with pervasive missingness. However, clinical applications demand predictive models that are both accurate and in...

📖 Read original article


70. DEFT: Data-Efficient Frequency-domain Top-k Sampling via Inverse Discrete Fourier Transform for Spatiotemporal Dynamical Systems Modeling ​

Author: Hengbo Xiao, Jiale Liu, Jiahao Song, Guannan He
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11019v1 Announce Type: new Abstract: Modeling spatiotemporal dynamical systems governed by partial differential equations (PDEs) poses two major challenges: it either requires expensive physics-based simulators that entail iterative numerical solving at high computational cost, or it depe...

📖 Read original article


71. Derivative Computation in PINNs: Automatic Differentiation, Finite Differences and Beyond ​

Author: Maciej J. Mikulski, Tadeusz Uhl
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, physics.comp-ph

arXiv:2608.11020v1 Announce Type: new Abstract: We systematically investigate finite-difference (FD) derivative computation in Physics-Informed Neural Networks (PINNs) as an alternative to automatic differentiation (AD). On three benchmark PDEs we show that, with a properly calibrated step size, FD ...

📖 Read original article


72. Mapping and Measuring the Behavioral Evolution of Large Language Models ​

Author: Dong Qiao, Chris Ding, Jicong Fan
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.11027v1 Announce Type: new Abstract: Benchmark leaderboards summarize how well a language model performs, but not how its behavior relates to that of other models or changes across generations. We characterize the output behavior of 32 models from six families using their responses to a s...

📖 Read original article


73. ReRound: Reconstructive Rounding to Resolve Midpoint Ambiguity in Calibration-Free LLM Quantization ​

Author: He-Yen Hsieh, H. T. Kung
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.11045v1 Announce Type: new Abstract: ReRound (Reconstructive Rounding) is a post-training quantization method that addresses the midpoint ambiguity inherent in standard round-to-nearest (RTN) schemes when quantizing weights near the centers of quantization intervals. Starting from a pretr...

📖 Read original article


74. Efficient Hypergradient Descent for Inverse Reinforcement Learning ​

Author: Nikita Sevriukov, Anna Barabanova, Uliana Gagarina, Karina Ivanova, Sofiia Kasaeva, Ilya Levin, Marina Sheshukova
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.11052v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) aims to recover a reward function under which the resulting policy reproduces the behavior observed in expert demonstrations. A natural approach is to formulate IRL as a bilevel optimization problem, in which the in...

📖 Read original article


75. Uncertainty-Aware Deep Learning for Genomics Applications: Insights from an Empirical Study ​

Author: Sepideh Saran, Mahsa Ghanbari, Uwe Ohler
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11054v1 Announce Type: new Abstract: Deep learning models have emerged as the standard computational tool for a wide range of applications in genomics. Yet, uncertainty quantification (UQ) -- and more specifically, the reliability of different uncertainty estimates in this domain -- has r...

📖 Read original article


76. Batch Size or Negatives? A Selection Rule for Memory-Constrained Recommender Training ​

Author: Artyom Sabitov, Daniil Volkov, Alexey Zaytsev
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11061v1 Announce Type: new Abstract: Large-scale neural recommender systems are typically trained with a softmax cross-entropy objective over the full item vocabulary. For a typical large number of possible items $K$, the final classification layer dominates memory, requiring $O(nK)$ logi...

📖 Read original article


77. Cross-View Feature Matching: Survey, Benchmarking, and Foundation-Model Perspectives ​

Author: Songlin Du, Xiaoyong Lu, Zeyu Wu, Xiaobo Lu, Guobao Xiao, Bin Fan, Jiayi Ma, Takeshi Ikenaga
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.11093v1 Announce Type: new Abstract: Cross-view feature matching aims to establish reliable correspondences across images with large viewpoint variations. Over the past decade, the field has evolved from task-specific models toward increasingly unified and generalizable correspondence mod...

📖 Read original article


78. Two-stage Odd Residual Flows for Mean-Preserving Probabilistic Time Series Forecasting ​

Author: Kiran Madhusudhanan, Christian Kl"otergens, Lars Schmidt-Thieme, Vijaya Krishna Yalavarthi
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11114v1 Announce Type: new Abstract: Probabilistic forecasting plays an essential role in risk-sensitive decision-making, particularly in long-horizon settings. However, existing approaches often face a fundamental trade-off between distributional flexibility and accurate mean prediction....

📖 Read original article


79. A Recommendation System Approach for Interference-Robust Sensor Subset Selection ​

Author: Kaan Buyukkalayci, Kyle Pak, Merve Karakas, Christina Fragouli
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11143v1 Announce Type: new Abstract: This paper develops a method for sensor-subset selection for tracking. Prior work showed that low-cost acoustic Received Signal Strength Indicator (RSSI) measurements can be used to recommend subsets of sensor nodes whose expensive sensing modalities, ...

📖 Read original article


80. DACRI: Decision-Aware Causal Intervention Ranking for Critical Supply Chains ​

Author: Shiqi Huang, Jiani He, Dingyan Shang, Yihua Xu, Jize Li, Yan Lyu, Lashimi Muraleedharan Nair
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11154v1 Announce Type: new Abstract: Detecting or attributing a supply-chain disruption is not the same as selecting the intervention that maximizes recoverable net value. We present CriticalSCM-Bench v1, a controlled synthetic benchmark with causal ground truth, paired factual/counterfac...

📖 Read original article


81. Hierarchical Empirical-Bayes Naive Bayes: Minimax Smoothing and Calibration with AODE Extension ​

Author: Nguyen Thai Anh, Truong Viet Vu, Tran Thien Thanh, Vo Nguyen Quoc Bao, Ngo Hoang Tu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11162v1 Announce Type: new Abstract: The Naive Bayes (NB) classifier remains a standard choice for categorical data, yet its widely used smoothing rules, such as Laplace, Lidstone, Krichevsky-Trofimov, and the $m$-estimate, all prescribe a fixed smoothing strength that ignores feature car...

📖 Read original article


82. Beyond a Bag of Features: Set-Level Instability in Sparse Autoencoders ​

Author: Nikolai Bolik, Lennart St"opler, Artur Andrzejak
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.11197v1 Announce Type: new Abstract: Shani et al. (2026) show that LLM representations broadly recover human category boundaries, while failing to reflect fine-grained typicality structure. Their analysis uses cosine similarity over dense model representations. We revisit their approach u...

📖 Read original article


83. Quantifying the noise sensitivity of the Wasserstein metric for images ​

Author: Erik Lager, Gilles Mordant, Amit Moscovich
Published: 8/12/2026, 4:00:00 AM
Categories: math.ST, cs.LG, eess.IV, stat.TH

arXiv:2510.01015v3 Announce Type: cross Abstract: Wasserstein metrics are increasingly adopted as similarity scores for images. We consider the sensitivity of Wasserstein metrics with respect to pixel-wise additive noise when the images are treated as discrete measures on the pixel grid. We derive f...

📖 Read original article


84. Optimized Sequential Testing for Binary Ensemble Classifiers ​

Author: Joseph Kalman, Amit Moscovich
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML

arXiv:2606.15237v1 Announce Type: cross Abstract: Ensemble classifiers are predictive models that combine the results of simpler base models, often by majority vote. A classic example is random forests, which combine the predictions of decision trees. Ensembles that use more base models can be more ...

📖 Read original article


85. Divergent Response Modes in Frontier Language Models Under Steering Pressure ​

Author: Ali Jalal-Kamali
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.06578v1 Announce Type: cross Abstract: Frontier language models are trained using distinct data, objectives, and safety pipelines. Whether these differences produce measurably different behaviors under explicit steering pressure remains underexplored. This study evaluates behavioral steer...

📖 Read original article


86. HyperShape: Hyperelasticity Across Diverse Shapes ​

Author: Leo Widmer, Sidaty El Hadramy, St'ephane Cotin, Philippe Claude Cattin
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CE, cs.LG

arXiv:2608.09938v1 Announce Type: cross Abstract: Hyperelastic deformations are highly sensitive to domain geometry and boundary conditions, making generalization across both a critical capability for neural operators applied to these problems. However, existing benchmarks for neural operators on hy...

📖 Read original article


87. When Chain-of-Thought Helps and When It Hurts: An Empirical Investigation of the Serial-Depth Bottleneck in LLM Reasoning ​

Author: Tughanbulut Kurtulush
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.09942v1 Announce Type: cross Abstract: It is widely assumed that chain-of-thought (CoT) prompting universally improves LLM reasoning. We investigate this through the conceptual framework of the H_dp bandwidth bound (Chen et al., 2024): although the formal bound binds only asymptotically (...

📖 Read original article


88. EweAcT: Ewe behaviour aligned to accelerometer data for activity monitoring in extensive grazing systems ​

Author: Lucile Riaboff (GenPhySE, INRAE), Ny Aina Andriamampandry (GenPhySE, GenPhySE), Jean-Fran\c{c}ois Bompa (GenPhySE, GenPhySE), Mathias Aletru (GenPhySE, GenPhySE), Christian Durand (UEF), S'ebastien Douls (UEF), Ga"etan Bonnafe (UEF), Morgane Costes-Thir'e (GenPhySE, GenPhySE), Guillaume Delosi`eres (GenPhySE, GenPhySE), Jean- Marc Mongrelet (GenPhySE, GenPhySE), Enzo Niro (GenPhySE, GenPhySE), N'emuel Tadi (GenPhySE, GenPhySE), S'everine Deretz (DEPT GA, UEF, INRAE), Sara Parisot (UEF), Margot Lamarque (UEF), Dominique Hazard (GenPhySE), Emilie Cobo (GenPhySE)
Published: 8/12/2026, 4:00:00 AM
Categories: cs.HC, cs.LG

arXiv:2608.09943v1 Announce Type: cross Abstract: Monitoring livestock behaviour under extensive conditions would provide valuable insights to assess animal adaption to environmental perturbations in agroecological systems (e.g., heat waves, parasitism, predator attacks). Animal behaviour can be mon...

📖 Read original article


89. An adaptive and evolvable deep reinforcement learning framework for weather prediction ​

Author: Qiang Wu, Han Li, Jianping Huang
Published: 8/12/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG

arXiv:2608.09948v1 Announce Type: cross Abstract: No single AI weather model excels at all variables, pressure levels, and lead times. Rather than building yet another architecture, we reframe the forecasting problem as one of coordination. Here we present Feitian Adaptive Ensemble Weather (FTAE-Wea...

📖 Read original article


90. AIFS-TC: A simple correction competitive with the operational frontier for tropical cyclone intensity forecasting ​

Author: Anna Allen, Wessel P. Bruinsma, Michael Maier-Gerber, Harrison Cook, Matthew Chantry, Richard E. Turner
Published: 8/12/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG

arXiv:2608.09959v1 Announce Type: cross Abstract: AI weather models are in the process of revolutionising weather forecasting. While these models have been shown to achieve superior performance to physics-based NWP in forecasting tropical cyclone (TC) tracks, they dramatically underestimate intensit...

📖 Read original article


91. Projected climate memory and inherited warm-tail risk in accelerated European summer warming ​

Author: Mauricio Herrera-Mar'in, Alex Godoy-Fa'undez, Diego Rivera
Published: 8/12/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG, stat.AP

arXiv:2608.09966v1 Announce Type: cross Abstract: European summer warming reflects interactions among background change, persistent ocean--land--circulation states, and same-season variability. We develop an empirical reduced-dynamics framework that decomposes regional summer indicators into inherit...

📖 Read original article


92. SPOTting the Future: Lookahead Explanations for Deep Reinforcement Learning ​

Author: Tamar Gozlan, Claudia V. Goldman
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.09967v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) agents achieve strong performance in complex environments, yet their decision-making processes remain difficult to interpret. We introduce SPOT (Sampling Policy Observation Tree), a novel model-agnostic, sampling-bas...

📖 Read original article


93. Do AI weather models miss extremes? ​

Author: Marvin Vincent Gabler, Roberto Molinaro, Niall Siegenheim, Henry Martin, Mark Frey, Niels Poulsen, Philipp Seitz, Olivier Lam
Published: 8/12/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.AI, cs.LG

arXiv:2608.09972v1 Announce Type: cross Abstract: First-generation AI weather models are often reported to underperform at extremes, mostly in reanalysis-based evaluations of deterministic regression systems. We verify eleven physical and AI forecast systems against European synoptic, solar, and rai...

📖 Read original article


94. MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment Analysis ​

Author: Yuhua Wen, Yingying Zhou, Qifei Li, Yingming Gao, Zhengqi Wen, Jianhua Tao, Ya Li
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.09986v1 Announce Type: cross Abstract: Most existing multimodal sentiment analysis approaches assume access to complete multimodal inputs. However, real-world applications frequently encounter incomplete or corrupted modalities, posing a critical challenge. Although several methods have b...

📖 Read original article


95. Knowledge-Guided 3D CT Generation: A Conditioning-Centric Taxonomy ​

Author: Francesca Pia Panaccione, Eugenio Lomurno, Matteo Matteucci
Published: 8/12/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG

arXiv:2608.09992v1 Announce Type: cross Abstract: Controllable generation guided by external knowledge is a key requirement in modern generative deep learning applications, enabling the synthesis of samples with explicit constraints on semantic content, structural properties, and variability. In 3D ...

📖 Read original article


96. Energy and Performance Benchmarking of Deep Learning Models for Breast Cancer Detection ​

Author: Samar Garrab, Ghada Achour
Published: 8/12/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG

arXiv:2608.09996v1 Announce Type: cross Abstract: Recent advances in machine learning have greatly improved breast cancer detection, enabling more accurate and timely diagnosis. Deep learning (DL) models show strong potential for medical image analysis; however, as their architectural complexity inc...

📖 Read original article


97. Towards Sustainable Artificial Intelligence: A Comprehensive Review and Comparative Analysis of Deep Learning Models' Carbon Footprint ​

Author: Samar Garrab, Sarra Boughriou, Manel BenSassi
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG, cs.SE

arXiv:2608.09998v1 Announce Type: cross Abstract: Artificial Intelligence (AI) and Machine Learning (ML) have become powerful tools for supporting and automating complex human tasks. Despite their benefits, growing attention has been directed toward their environmental implications, primarily due to...

📖 Read original article


98. Do LLM Recommenders Know When They're Hallucinating? Auditing Confidence Calibration in Catalog Faithfulness ​

Author: Srijith Ravikumar
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IR, cs.CL, cs.LG

arXiv:2608.10008v1 Announce Type: cross Abstract: LLM recommenders for top-$K$ item suggestion regularly emit titles outside the target catalog. Prior audits measure this as a binary out-of-domain rate; none ask whether the model knew it was hallucinating. We jointly audit hallucination rate (OOD@10...

📖 Read original article


99. HIPNO: Symmetry-Aware Physics-Informed Neural Operators for Noninvasive Hemodynamic Inference ​

Author: Yunbei Pan, Jiahang Sha, Simon A. Lee, Maxime Cannesson, Wei Wang, Jeffrey N. Chiang
Published: 8/12/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG, eess.SP

arXiv:2608.10011v1 Announce Type: cross Abstract: Continuous hemodynamic monitoring guides treatment decisions in surgery and intensive care. However, gold-standard signals are only measured in severe cases due to risks associated with invasive measurement. In this work, we introduce HIPNO (Hemodyna...

📖 Read original article


100. Deep Learning-Based Statistical Downscaling of Sea Surface Temperature Using a Residual Corrective Neural Network ​

Author: Onkar Jadhav, Tim French, Ivica Janekovic, Nicole L. Jones, Matthew Rayson
Published: 8/12/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG

arXiv:2608.10022v1 Announce Type: cross Abstract: The large-scale oceanic and atmospheric forecasts provided by global climate models typically lack sufficient resolution to accurately capture the response of the coastal ocean to atmospheric forcing and coastal circulation that drive fine-scale SST ...

📖 Read original article


101. Navigating the Proximity-Safety Balance: Constraint Decomposition for Human Following in Pedestrian Crowds ​

Author: Shiting Gong, Jianpeng Yao, Jinfeng Wang, Marco Pavone, Jiachen Li
Published: 8/12/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, cs.SY, eess.SY

arXiv:2608.10056v1 Announce Type: cross Abstract: Following a target human in crowded environments involves an inherent conflict between staying close to the target and navigating safely among surrounding pedestrians and obstacles. This conflict becomes more severe in dense scenarios, where aggressi...

📖 Read original article


102. Status Association Does Not Reliably Predict Decision Leakage ​

Author: Abdullah X
Published: 8/12/2026, 4:00:00 AM
Categories: stat.AP, cs.AI, cs.CY, cs.LG

arXiv:2608.10089v1 Announce Type: cross Abstract: Bias evaluations often move too quickly from evidence that a model encodes a social association to claims that the same association will alter consequential decisions. We test whether that inference is warranted using Chilean surnames as controlled s...

📖 Read original article


103. CHORUS: Complementary Experts for High-Coverage Testbench Stimulus Generation ​

Author: Hejia Zhang, Sheng Lu, Zhongming Yu, Chia-Tung Ho, Brucek Khailany, Jishen Zhao
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.10090v1 Announce Type: cross Abstract: Large language models (LLMs) have advanced code generation, where executable feedback provides a more reliable learning signal than textual imitation alone. Hardware verification is an important application of code generation and accounts for a subst...

📖 Read original article


104. Deciding When to Switch: E-Processes for Adaptive Minimax Training for Generative Adversarial Nets ​

Author: Hyunjoo Kim, Sicheng Wu, Agastya Venkatraman, Guang Lin, Sehwan Kim
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.10096v1 Announce Type: cross Abstract: Modern data science increasingly gives rise to hypothesis-testing problems that are not naturally formulated in terms of parameters within prespecified statistical models. One important example is the dynamic evaluation of optimization algorithms, wh...

📖 Read original article


105. P3CA: Encoder-Agnostic Interpretation of Vision Foundation Model Embeddings via Spatial Probing ​

Author: Amoon Jamzad, Dilakshan Srikanthan, Faranak Akbarifar, Nooshin Maghsoodi, Parvin Mousavi
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.10131v1 Announce Type: cross Abstract: Vision foundation models are increasingly used as reusable encoders in medical image computing, yet their high-dimensional spatial embeddings are difficult to inspect beyond downstream task performance or global dimensionality reduction. We propose p...

📖 Read original article


106. The Parser Already Knows: Lightweight Bias Correction in Constrained Decoding ​

Author: I\c{s}{\i}l "Ozg"u, Yaoxuan Wu, Guy Van den Broeck, Miryung Kim
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.10137v1 Announce Type: cross Abstract: Grammar Constrained Decoding (GCD) forces Language Models (LMs) to produce syntactically valid outputs by masking out non-conforming tokens at each step. However, rigid masking distorts the model's underlying probability distribution, often biasing g...

📖 Read original article


107. More Accurate, Less Human: Gestalt Grouping in Vision Models ​

Author: Sudhanva Manjunath Athreya, Sai Phani Kumar Malladi
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.10195v1 Announce Type: cross Abstract: Human vision organizes what it sees into wholes: same-colored points group into series, similar marks cohere into categories, and shapes complete into recognizable objects. These are the Gestalt operations that visualization design builds on. Whether...

📖 Read original article


108. FACT: Failure-Aware Causal Training for World-Action Models ​

Author: Quanquan Peng, Yutong Liang, Rui Yan, Nicklas Hansen, Xiaolong Wang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2608.10232v1 Announce Type: cross Abstract: Recent world-action models (WAMs) show that co-training policies with future prediction can provide physical priors for action generation. Building on the future-prediction ability of video models, many WAMs generate future videos and recover actions...

📖 Read original article


109. The Kuramoto Neural Operator: Learning to Solve PDEs via Coupled Oscillator Dynamics ​

Author: Petr Badolia, Leonid Obukhov, Dmitry Bylinkin, Aleksandr Beznosikov
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CE, cs.LG

arXiv:2608.10234v1 Announce Type: cross Abstract: Operator learning is a rapidly advancing area of computational science. It is particularly well suited to problems where a partial differential equation (PDE) must be solved repeatedly under varying physical configurations. Most existing architecture...

📖 Read original article


110. Sequential Modality Dropout for Robust Multi-Modal Sequential Recommendation ​

Author: Guanqun Yang, Wenlong Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, cs.MM

arXiv:2608.10240v1 Announce Type: cross Abstract: Multi-modal sequential recommenders assume every item carries every modality, but real product catalogs often miss images or text, and a model trained on complete data loses much of its recommendation accuracy when a modality is unavailable at servin...

📖 Read original article


111. A Graph Neural Network--Guided Genetic Algorithm for Physical Internet Supply Chain Optimization under Cost Uncertainty ​

Author: Faezeh Ardali, Gerald M. Knapp
Published: 8/12/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2608.10245v1 Announce Type: cross Abstract: Inventory and distribution planning in Physical Internet networks requires coordinating factory-hub assignments, factory supply, lateral transshipment among collaborative hubs, retailer deliveries, and shortages. The problem combines discrete assignm...

📖 Read original article


112. DualSpectralCF: Training-Free Sign-Aware Spectral Collaborative Filtering ​

Author: Guanqun Yang, Tong Qi, Xiaoxue Han
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, cs.SI

arXiv:2608.10247v1 Announce Type: cross Abstract: Real-world recommendation platforms routinely collect explicit negative feedback such as 1-star reviews, hate-button clicks, distrust between users, and very-low watch-ratio videos. Learned sign-aware recommenders exploit this signal for clear accura...

📖 Read original article


113. Off-Axis, On Purpose: Where a Transformer Computes Concepts and Why it Does So ​

Author: Mark Oskin
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.10251v1 Announce Type: cross Abstract: A transformer's answer lives on one axis: the direction its unembedding reads. Its intermediate states largely do not, and that off-axis position is usually treated as an obstacle to interpretation. We show it is functional. A 12-layer model computes...

📖 Read original article


114. BreastMammo and DenseMammo: Benchmarks for Mammography Domain Generalization ​

Author: Hongyi Pan, Gorkem Durak, Halil Ertugrul Aktas, Andrea Mia Bejar, Mustafa Ege Seker, Nebile Alibeyoglu, Rumeysa Guclu, Rana Gunoz Comert Bozkurt, Sibel Ozkan Gurdal, Neslihan Cabioglu, Beyza Ozcinar, Ravza Yilmaz, Vahit Ozmen, Erkin Aribal, Sukru Mehmet Erturk, Yalda Zafari, Mohamed Mabrok, Kayhan Batmanghelich, Mohammad Yaqub, Ziyue Xu, Ulas Bagci
Published: 8/12/2026, 4:00:00 AM
Categories: eess.IV, cs.LG

arXiv:2608.10271v1 Announce Type: cross Abstract: Breast density classification is a critical component of breast cancer risk assessment, yet AI models often struggle to generalize across clinical sites due to vendor-specific acquisition styles. In this work, we introduce two new datasets, BreastMam...

📖 Read original article


115. Stochastic Emulation of a Fully Coupled Preindustrial E3SMv3 Simulation ​

Author: Elynn Wu, James P. C. Duncan, Troy Arcomano, Jeremy McGibbon, Oliver Watt-Meyer, Christopher S. Bretherton, Naser Mahfouz, Claudia Tebaldi, Luke Van Roekel, Andrew Roberts, Wuyin Lin, Finn Rebassoo, Jean-Christophe Golaz, Peter M. Caldwell
Published: 8/12/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG

arXiv:2608.10277v1 Announce Type: cross Abstract: We present a stochastic coupled emulator of E3SM version 3, built on the SamudrACE framework, which couples an atmosphere emulator (ACE2) with a full-depth ocean emulator (Samudra). We replace the deterministic atmosphere emulator with its stochastic...

📖 Read original article


116. SeFaR: Semantic Feature-aware Robustness Testing of Deep Neural Networks ​

Author: Nusrat Jahan Mozumder, Divya Gopinath, Corina Pasareanu, Matthew Dwyer
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.10289v1 Announce Type: cross Abstract: Deep neural networks are increasingly deployed in safety-critical domains as perception modules, where failures are often caused due to rare and under-represented scenarios. This necessitates the need to evaluate the semantic robustness of perception...

📖 Read original article


117. MRIComp4Flow: Compression of 3D Brain MRI for Training Multi-Modal Generative Models ​

Author: Lisa K. Fischer, Mykhailo Riabets, Daniel Rueckert, Benedikt Wiestler, Anke Meyer-Baese, Sandeep Nagar
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.10291v1 Announce Type: cross Abstract: Large-scale multi-modal MRI datasets impose substantial storage and I/O costs, limiting the training of 3D generative models on commodity infrastructure. While lossy compression is known to preserve accuracy for discriminative segmentation networks, ...

📖 Read original article


118. Logit-Boundary Geometric Belief Interfaces and Sparse Sheaf-Enclave Protocols: A Self-Contained Substrate for Secure Network Electronic Health Record (EHR) Interoperability ​

Author: Alvin Spivey, Yu Huang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG

arXiv:2608.10300v2 Announce Type: cross Abstract: Electronic health-record interoperability is a boundary problem: legacy systems, generative models, terminology services, identity systems, and human reviewers may each expose rich internal states, while operational exchange requires a narrow shared ...

📖 Read original article


119. UniMod: Enhancing Multi-Modal Medical Diagnosis through Cross-Modality and Within-Modality Alignment ​

Author: Zijian Gu, Weikai Lin, Shuang Zhou, Zihan Chen, Song Wang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.MM

arXiv:2608.10316v1 Announce Type: cross Abstract: Multi-modal learning combining medical images and clinical text is promising for disease diagnosis. However, standard multi-modal training leads to shortcut learning: models exploit the easier modality (e.g., diagnostic cues in text) while neglecting...

📖 Read original article


120. Topological Feasibility Guarantees for Differentiable Predictive Control ​

Author: Guangyu Wu, J'an Drgo\v{n}a
Published: 8/12/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2608.10332v1 Announce Type: cross Abstract: Differentiable predictive control (DPC), a self-supervised learning approach for approximating explicit model predictive control (MPC) policies, offers significant computational advantages over online optimization-based MPC. However, feasibility guar...

📖 Read original article


121. On the Importance of Geometric Nonlinearity and Temperature-Dependent Properties in Multi-Material Thermo-Mechanical Topology Optimization ​

Author: Shirin Hosseinmardi, Xiangyu Sun, Ramin Bostanabad
Published: 8/12/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.CE, cs.LG

arXiv:2608.10344v1 Announce Type: cross Abstract: Thermo-mechanical compliant devices are commonly designed with small-strain linear elasticity and temperature-independent material properties, even though they might operate hundreds of kelvin above ambient where both assumptions are questionable. In...

📖 Read original article


122. Beyond Detection Accuracy: Measuring Explanation Cost, Stability, and Utility for Resource-Aware IoT Intrusion Detection ​

Author: Abdurrahman Tolay
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.NI

arXiv:2608.10349v1 Announce Type: cross Abstract: Machine-learning intrusion-detection studies commonly emphasize predictive accuracy while treating explanation generation as a computationally free post-processing step. This study jointly evaluates predictive effectiveness, explanation cost, local e...

📖 Read original article


123. Efficient Weak-Entropy PINN for Solving Hyperbolic Conservation Laws ​

Author: Qi Gao, Kuang Huang, Xuan Di
Published: 8/12/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA

arXiv:2608.10389v1 Announce Type: cross Abstract: In recent years, neural networks have significantly advanced numerical solutions of partial differential equations (PDEs). However, solving PDEs with discontinuous solutions, such as hyperbolic conservation laws, remains challenging for neural networ...

📖 Read original article


124. Automatic Field-of-View Adjustment for a View-Expansive Microscope via LSTM-Based Gaze and Pipette Motion Interpretation ​

Author: Kenta Yokoe, Takuya Hara, Tadayoshi Aoyama
Published: 8/12/2026, 4:00:00 AM
Categories: cs.HC, cs.LG, cs.RO, eess.IV

arXiv:2608.10401v1 Announce Type: cross Abstract: Intracytoplasmic sperm injection (ICSI) operators frequently adjust the field-of-view (FOV) during procedures, which interrupts workflow and increases procedure time. Conventional microscopes require manual objective lens switching and illumination a...

📖 Read original article


125. Post-Calibration Reliability Reranking of Relevance Decisions via Label-wise Monotone Projection ​

Author: Inwoo Tae, Yongjae Lee
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.10406v1 Announce Type: cross Abstract: Web search, product search, and question-answering retrieval systems often assign a relevance label and confidence score to each query-candidate pair. The relevance label describes how well a page, product, or passage matches the query, while the con...

📖 Read original article


126. How Robust Are LLMs to Vietnamese Dialects? ​

Author: Minh Tran, Trinh Chau, Thanh-Nhan Le, Nam Tran, Luan Thanh Nguyen, Cuong Dang, Duc Hoang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.10414v1 Announce Type: cross Abstract: Large Language Models (LLMs) are typically evaluated on standard written Vietnamese, yet everyday communication frequently involves regional dialects that preserve meaning but differ in surface form. Existing Vietnamese dialect work largely addresses...

📖 Read original article


127. Riemann GeoResolver: A Non-Euclidean Attention Framework from Euclidean Resolver to Hyperbolic-Spherical Geometry ​

Author: Liangchen Ge
Published: 8/12/2026, 4:00:00 AM
Categories: cs.DS, cs.AI, cs.CL, cs.LG

arXiv:2608.10416v1 Announce Type: cross Abstract: We present a theoretical foundation for inverse-distance attention, from its Euclidean prototype (Resolver) to its non-Euclidean realization (Riemann GeoResolver). The Euclidean part establishes three core theorems: (1) circuit separation---IDA achie...

📖 Read original article


128. A lower bound for stepsize-based acceleration of gradient descent ​

Author: Jianhao Ma, Yuxin Chen
Published: 8/12/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ML

arXiv:2608.10418v1 Announce Type: cross Abstract: Recent work has shown that, for smooth convex optimization, plain gradient descent can be accelerated from its textbook convergence rate of $O(T^{-1})$ (where $T$ denotes the number of iterations) to $O\big(T^{-\log_2(1+\sqrt{2})}\big)$ using careful...

📖 Read original article


129. Recovering Wasted Compute in Autoresearch Agents ​

Author: Au Kwok Chun, Abhigyan Acherjee, Amrutha Rao, Zaiqian Chen, Kazem Meidani, C. Bayan Bruss, Micah Goldblum
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.10424v1 Announce Type: cross Abstract: A slew of recent works develop agents for solving research problems end-to-end, a paradigm increasingly referred to as autoresearch. Such agents have inspired large industry investment, motivated by their potential to automate time-consuming human la...

📖 Read original article


130. Quantum Incremental Learning with Mixed State Prototypes ​

Author: Yu Wu, Qianli Zhou, Xinyang Deng, Wen Jiang, Kang Hao Cheong, Witold Pedrycz
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.10464v1 Announce Type: cross Abstract: Incremental learning models are required to learn new classes sequentially without catastrophic forgetting, while operating under parameter and memory constraints. In the Noisy Intermediate-Scale Quantum (NISQ) era, although quantum neural networks o...

📖 Read original article


131. Stay or Stray - A Dynamical Systems Viewpoint of Popularity Bias ​

Author: Sarvesh Shashidhar, Lankireddy Prabhat, Arpit Agarwal, D. Manjunath, Karan Bhukar, Tanmay Khandelwal
Published: 8/12/2026, 4:00:00 AM
Categories: cs.HC, cs.LG

arXiv:2608.10474v1 Announce Type: cross Abstract: Popularity bias in recommendation systems arises when a majority user class generates disproportionate interaction data, causing the system to increasingly favour it while degrading recommendation quality for niche users. While extensive empirical ev...

📖 Read original article


132. Multi-Granular Rationale-Guided Molecular LLM for Property Prediction ​

Author: Junwoo Park, Minyoung Shin, Cheol Soon Lee, Sujee Lee
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.10480v1 Announce Type: cross Abstract: Large language models (LLMs) are widely applied across chemical tasks, such as molecular property prediction, which underpins drug discovery. Molecular LLMs represent a molecule through several modalities, notably a 1D SMILES sequence or a 2D molecul...

📖 Read original article


133. CARB: A Characterization-Guided Framework for CNN Inference Cost Prediction and Deployment Screening ​

Author: Linh Nguyen, Zhixin Pan
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, cs.PF

arXiv:2608.10506v1 Announce Type: cross Abstract: Accurate pre-deployment estimation of CNN inference cost--energy, latency, and peak memory--is increasingly critical as models are deployed on resource-constrained GPU platforms. Existing approaches rely on FLOPs, latency measurements, or single-devi...

📖 Read original article


Author: Xiaoxuan Gao, Rentao Gu, Yingchun Wang, Xinyi Liu, Junshi Gao, Yuefeng Ji
Published: 8/12/2026, 4:00:00 AM
Categories: cs.NI, cs.LG, physics.data-an, physics.optics

arXiv:2608.10517v1 Announce Type: cross Abstract: Accurate physical-layer modeling is increasingly essential for reliable ultra-wideband operation and capacity optimization, especially under the intensified inter-channel stimulated Raman scattering (ISRS) effect. This paper proposes the link-adaptiv...

📖 Read original article


135. When Do Anchor-Based Pointwise LLM Rerankers Help? Retriever Quality, Statistical Scope, and Anchor Design ​

Author: Utshab Kumar Ghosh, Shubham Chatterjee
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.10528v2 Announce Type: cross Abstract: Anchor-based pointwise LLM reranking scores each candidate against a shared reference passage to recover cross-document context at pointwise cost. We study when this actually helps, using GCCP/PAGC as a representative method. Our study is reproductio...

📖 Read original article


136. Benchmarking LLM-Guided Control-Plane Policies for Backend Fault Isolation in HAProxy ​

Author: Aman Chauhan, Vishnu Pendyala
Published: 8/12/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2608.10532v1 Announce Type: cross Abstract: Static load balancers cannot mitigate a backend that is degraded rather than down: round-robin and least-connections keep routing traffic to a server returning HTTP 500s until an operator intervenes. We ask whether a Large Language Model can replace ...

📖 Read original article


137. Measuring Semantic Abstractness of SAE Features via Nonlocality ​

Author: Chuqiao Lin, Shivaji Sondhi, Xiao-Liang Qi
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.10537v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have helped uncover mechanistic explanations for LLM behaviours such as reasoning, jailbreaking etc., via understanding the corresponding task-relevant and causally effective features. To evaluate such mechanistic explanati...

📖 Read original article


138. Flow Straight to Reality: Perceptually Consistent Flow Matching for Efficient Image Restoration ​

Author: Sangwoo Jo, Donggeun Ko, Jayeon Kang, Youngsang Kwak, Jaehwa Kwak, Sungjoon Choi
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.10544v1 Announce Type: cross Abstract: Image restoration is fundamentally constrained by the tradeoff between distortion and perception: minimizing pixel-wise error yields over-smoothed results, whereas optimizing for perceptual realism often introduces structural deviations. Recent appro...

📖 Read original article


139. Iterative Erasure Count Is Not an Affine-Invariant Concept Dimension ​

Author: Tingan Jin, Shuhang Dong, Haosong Li, Chung-Hsien Chou
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.CV, cs.LG

arXiv:2608.10566v1 Announce Type: cross Abstract: How many directions does a neural representation use to encode a concept? A common answer repeatedly erases probe directions and reports the stopping count or cumulative removed rank. We show that both quantities can change under an information-prese...

📖 Read original article


140. BooST: Bridging Semantics and Motions for Efficient Skill Transfer ​

Author: Jusuk Lee, Daesol Cho, Jonghun Shin, Seungyeon Yoo, Jonghae Park, Taekbeom Lee, H. Jin Kim
Published: 8/12/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2608.10600v1 Announce Type: cross Abstract: Skill abstraction---the process of learning reusable and temporally extended behaviors---has emerged as a key paradigm for improving sample efficiency and generalization in robot learning. For efficient skill transfer to real robots, learned skills m...

📖 Read original article


141. InSight-doc: Agentic Visual Perception for Long-Document Understanding ​

Author: Kaican Li, Weiyan Xie, Lewei Yao, Jiannan Wu, Lanqing Hong, Yongxiang Huang, Nevin L. Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2608.10628v1 Announce Type: cross Abstract: Long-document understanding often requires reasoning over many visually rich pages, making inference costly and prone to context rot. In this work, we propose InSight-doc, an agentic visual perception framework that treats visual resolution as an ada...

📖 Read original article


142. Retrieval-Augmented Vision Foundation Models for Robust Leukemia Cell Classification across Multiple Microscopy Datasets ​

Author: Carlos Zamora, Hiram Zuniga, Ulises Orozco-Rosas, Kenia Picos
Published: 8/12/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2608.10657v1 Announce Type: cross Abstract: Leukemia cell image classification is challenged by real-world domain shifts from acquisition, staining, illumination, and site protocols, causing single-dataset models to generalize poorly in real clinical scenarios. This work presents a robust fram...

📖 Read original article


143. Rule of Thumb: Explaining Artificial Intelligence Systems using Partial Information ​

Author: Kaivalya Rawal, Daria Onitiu, Brent Mittelstadt, Sandra Wachter, Chris Russell
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, stat.ML

arXiv:2608.10766v1 Announce Type: cross Abstract: Explainable Artificial Intelligence (XAI) seeks to explain how an Artificial Intelligence (AI) system arrived at a particular decision. We propose ''Rule of Thumb'' (RoT) explanations, a new approach to XAI based upon a novel formulation that identif...

📖 Read original article


144. ChemWorld: Programmable Chemical Worlds for Controlled and Replayable Agent Experimentation ​

Author: Jiangjie Qiu, Yijun Li, Xiaonan Wang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.10792v1 Announce Type: cross Abstract: Autonomous chemistry increasingly depends on environments in which agents can repeatedly act, observe, and adapt.Physical laboratories provide essential real-material evidence but are costly to repeat and difficult to use for tightly matched interven...

📖 Read original article


145. Beyond Fixed Luminance: Towards Panchromatic and Orthochromatic Image Colorization ​

Author: Swarnim Maheshwari, Syed Imam Ali, Vineeth N. Balasubramanian
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.10798v2 Announce Type: cross Abstract: Most image colorization systems operate in $Lab$ space by predicting chroma ($ab$) while preserving an input-derived luminance channel ($L$). While effective on standard benchmarks, this fixed-luminance design restricts brightness changes and becomes...

📖 Read original article


146. BPG: Balancing Plasticity and Generalization for Domain Incremental Learning ​

Author: Qiang Wang, Songlin Dong, Shaokun Wang, Jizhou Han, Xiang Song, Chenhao Ding, Yuhang He, Yihong Gong
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.10804v1 Announce Type: cross Abstract: Deep neural networks excel in various tasks but struggle to generalize across evolving data distributions, leading to significant performance degradation under domain shifts. Domain incremental learning (DIL) addresses this challenge by enabling mode...

📖 Read original article


147. UniProbe: A Learnable Token-Level Hallucination Detector for Large VLMs using Multi-Structural Internal Representations ​

Author: Dvir Samuel, Guy Bar-Shalom, Fabrizio Frasca, Ethan Fetaya, Yftah Ziser, Gal Chechik, Haggai Maron
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.10835v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) achieve impressive visual reasoning and dialogue capabilities, yet frequently hallucinate content unsupported by the visual input. Effective mitigation requires token-level localization, enabling targeted interven...

📖 Read original article


148. Spectral Embeddings of Degree-$\alpha$ Laplacians in Random Dot Product Graphs ​

Author: John Park, Ning Hao
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.10845v1 Announce Type: cross Abstract: Spectral clustering methods for network data are commonly based on a few matrix representations, such as the adjacency matrix and the symmetric Laplacian. We study a continuum of degree-normalized spectral embeddings that includes these commonly used...

📖 Read original article


149. Self-Normalized Inference for Constant-Stepsize Temporal-Difference Learning under Markovian Sampling ​

Author: Min Zeng, Yichen Zhang, Xiaofeng Shao
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.10896v1 Announce Type: cross Abstract: Constant-stepsize temporal-difference (TD) learning is attractive for policy evaluation, but inference from a single Markov trajectory must account for serial dependence and a stepsize-dependent stationary target. For fixed-stepsize linear TD, we est...

📖 Read original article


150. VIDS-Seg: Towards Reliable Uncertainty Quantification in Pediatric Cardiac Ultrasound Segmentation ​

Author: Paul Fischer, Ece Ozkan
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.10903v1 Announce Type: cross Abstract: Reliable clinical deployment of machine learning requires models that know when they are likely to fail, particularly for subgroups underrepresented in training data. A common case is pediatric care, where models trained on adult cohorts can silently...

📖 Read original article


151. Threshold Structure of Optimal Policies in Restart POMDPs ​

Author: Konstantin Avrachenkov, Alexey Piunovskiy, Yi Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.PR

arXiv:2608.10936v1 Announce Type: cross Abstract: We study a Restart POMDP (Partially Observable Markov Decision Process) on a general Borel state space, where the controller either lets the hidden state evolve unobserved or restarts the system and observes the new state. Exploiting a sufficient-sta...

📖 Read original article


152. Information Bottleneck under Perfect Privacy ​

Author: Junle Zhong, Mohamad Assaad, Sreejith Sreekumar
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT

arXiv:2608.11003v1 Announce Type: cross Abstract: In this work, we study the information bottleneck under perfect privacy, with particular emphasis on the active-rate regime, where the representation-rate constraint is binding and directly limits the achievable utility. The goal is to construct a re...

📖 Read original article


153. Gromov-Wasserstein Quantization and Clustering: Structure, Rates, and Algorithms ​

Author: Florian Beier, Stephan Eckstein
Published: 8/12/2026, 4:00:00 AM
Categories: math.OC, cs.CG, cs.LG, math.PR

arXiv:2608.11016v1 Announce Type: cross Abstract: Clustering is a fundamental class of data analysis techniques with the most important representatives being centroid-based methods like $k$-means. Such methods are strongly connected to quantization problems, which aim to approximate general probabil...

📖 Read original article


154. SCOUT: Symmetric Consensus Outlier Detection for Failure Localization in LLM Pre-Training ​

Author: Zhuang Wang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.11034v1 Announce Type: cross Abstract: In LLM pre-training, synchronization propagates rank-local stalls, slowdowns, and numerical errors into job-wide symptoms, obscuring their origin. Existing diagnosis often relies on in-process monitors that cannot report after the trainer blocks or t...

📖 Read original article


155. V-FiLLM: Verified Financial LLM Reasoning Benchmark ​

Author: Alicia Larsen, Victoire Laurent, Aulia Kharis Rakhamsari, Lara Turgut, Nino Antulov-Fantulin
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.CE, cs.LG

arXiv:2608.11047v1 Announce Type: cross Abstract: While existing benchmarks have made substantial progress in evaluating LLMs across STEM domains, financial reasoning over structured data remains comparatively less explored. We introduce V-FiLLM, a framework that generates financial reasoning benchm...

📖 Read original article


156. A Systematic Sample Size Analysis of ML-Based Path Loss Prediction for LPWAN ​

Author: Robert Bitterling, Christian Nettersheim, J"orn Hees, Michael Rademacher
Published: 8/12/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2608.11083v1 Announce Type: cross Abstract: Low Power Wide Area Networks like LoRa are increasingly deployed for smart city applications, requiring accurate path loss prediction for effective network planning. Traditional (empirical) propagation models often exhibit limited accuracy in these s...

📖 Read original article


157. Why Does CLAUDE.md Keep Growing? Catastrophic Remembering in Agentic Coding ​

Author: Kushal Chakrabarti
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE

arXiv:2608.11095v1 Announce Type: cross Abstract: Agentic coding READMEs like CLAUDE.md grow without bound in real repositories, stopping only when the repository retires or someone rewrites the file wholesale. We trace this to imperfect recall: appending an instruction is always cheap, but once an ...

📖 Read original article


Author: Vladimir Iglovikov
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.11123v1 Announce Type: cross Abstract: Augmentation can corrupt a training example when an image and its annotations receive different random changes. A crop must use the same coordinates for the image, mask, boxes, keypoints, stereo views, video frames, or volume. Code paths that choose ...

📖 Read original article


159. Scheduling Mixed RL Rollouts Beyond Prefix Locality ​

Author: Zetao Hong, Song Yuan, Yuanhao Ding, Yibo Zhu, Daxin Jiang, Zhibin Wang, Chen Tian
Published: 8/12/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.11152v1 Announce Type: cross Abstract: Modern reinforcement learning (RL) post-training pipelines for large language models (LLMs) increasingly combine rollout workloads across multiple domains and feedback paradigms. Prefix-aware routing improves inference efficiency through cache reuse ...

📖 Read original article


160. Conditional Independence Tests for Constraint-Based Causal Discovery: A Survey ​

Author: Pavel Averin, Theodoros Moysiadis, Ioannis Katakis
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.11156v1 Announce Type: cross Abstract: Conditional Independence (CI) tests are the statistical engine of constraint-based causal discovery: in algorithms such as PC (Peter-Clark) and FCI (Fast Causal Inference), skeleton pruning and key orientations follow directly from CI decisions. This...

📖 Read original article


161. MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment ​

Author: Changhao Xiang, Shangyu Xing, Zhen Wu, Jianbing Zhang, Xinyu Dai
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2608.11167v1 Announce Type: cross Abstract: Existing Multimodal Large Language Models (MLLMs) predominantly rely on image-text pairs for modality alignment pretraining, mapping global image representations to long textual descriptions. However, this image-level alignment suffers from referenti...

📖 Read original article


162. A Quantum Roadmap for Softmax Attention: Exact Born-Rule Analogs for Softmax Attention on the Probability Simplex ​

Author: Eric A. F. Reinhardt, Adam J. Hauser
Published: 8/12/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.11173v1 Announce Type: cross Abstract: The attention mechanism forms the foundation of many modern AI models such as the Transformer. In one subclass of problems where attention is used, inputs and outputs are bound to the probability simplex so that all outputs sum to one. In this settin...

📖 Read original article


163. How to Verify Consistency of Probabilistic Claims ​

Author: Orr Paradise, Oliver Richardson, Yoshua Bengio, Shafi Goldwasser
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CC, cs.AI, cs.LG

arXiv:2608.11181v1 Announce Type: cross Abstract: When a probabilistic predictor answers many conditional-probability queries, are its answers self-consistent, and can this be verified in polynomial time? This problem is of interest for AI safety, where safety is derived from honesty about probabili...

📖 Read original article


164. ConVAWG: A Retrieval-Grounded Framework for Controlled Synthetic Dialogue Generation in Violence Against Women and Girls ​

Author: Chen Lyu, Xingwei Tan, Simon Cullen, Shelley Wilson, Lois Arthurs, Arshad Jhumka, Gabriele Pergola
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.11200v1 Announce Type: cross Abstract: Synthetic dialogue generation offers a way to study conversational dynamics in sensitive domains where real data are difficult to access, release, or annotate. The underlying abuse may occur online or offline: threats and coercion can appear directly...

📖 Read original article


165. Weighted Sequential Bayesian Inference for Non-Stationary Linear Contextual Bandits ​

Author: Nicklas Werge, Yi-Shan Wu, Abdullah Akg"ul, Melih Kandemir
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2307.03587v4 Announce Type: replace Abstract: In non-stationary linear contextual bandits, existing efficient algorithms typically rely on the Weighted Regularized Least-Squares (WRLS) estimator. Because WRLS only provides point estimates, previous methods typically construct surrogate distrib...

📖 Read original article


166. Convergence of Sign-based Random Reshuffling Algorithms for Nonconvex Optimization ​

Author: Zhen Qin, Zhishuai Liu, Pan Xu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, math.OC, stat.ML

arXiv:2310.15976v4 Announce Type: replace Abstract: signSGD is attractive in nonconvex optimization because it communicates sign-valued rather than full-precision gradients. Several standard analyses assume independent stochastic-gradient samples, whereas a common finite-sum implementation reshuffle...

📖 Read original article


167. High-Dimensional Calibration from Swap Regret ​

Author: Maxwell Fishelson, Noah Golowich, Mehryar Mohri, Jon Schneider
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, cs.GT, stat.ML

arXiv:2505.21460v2 Announce Type: replace Abstract: We study online calibration of multi-dimensional forecasts over an arbitrary convex set $P \subset \mathbb{R}^d$ relative to an arbitrary norm $|\cdot|$. We connect this to external regret minimization for online linear optimization (OLO): if one c...

📖 Read original article


168. Demystifying Adversarial Robustness in Diffusion Models: Compression, Randomness, and Geometry ​

Author: Liu Yuezhang, Xue-Xin Wei
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2505.22839v2 Announce Type: replace Abstract: Recent studies suggest that diffusion models significantly improve the empirical adversarial robustness of deep neural network models. While intuitive explanations have been proposed, the mechanisms underlying diffusion-based robustness remain larg...

📖 Read original article


169. TS-Mob: Social and Geographical-Aware Time Series Foundation-Model Framework for Human Mobility Prediction ​

Author: Massimiliano Luca, Ciro Beneduce, Bruno Lepri
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2507.00945v3 Announce Type: replace Abstract: Short-term forecasting of aggregated human mobility flows supports urban planning, intelligent transportation systems, and emergency response, yet existing models often require substantial mobility history and learn spatial structure implicitly thr...

📖 Read original article


170. DQS: A Low-Budget Query Strategy for Enhancing Unsupervised Data-driven Anomaly Detection Approaches ​

Author: Lucas Correia, Jan-Christoph Goos, Thomas B"ack, Anna V. Kononova
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.05663v4 Announce Type: replace Abstract: Truly unsupervised approaches for time series anomaly detection are rare in the literature. Those that exist suffer from a poorly set threshold, which hampers detection performance, while others, despite claiming to be unsupervised, need to be cali...

📖 Read original article


171. GLAM: Efficient Continual Learning at Scale via Grouped LoRA Adapter Merging ​

Author: Irene Testa, Luigi Quarantiello, Eric Nuertey Coleman, Samrat Mukherjee, Julio Hurtado, Vincenzo Lomonaco
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.13211v4 Announce Type: replace Abstract: The ability to learn continuously over time remains a major challenge for modern machine learning systems, even in the era of Foundation Models. While the rich representations learned by large pre-trained models can partially mitigate catastrophic ...

📖 Read original article


172. URS: A Unified Neural Routing Solver for Cross-Problem Zero-Shot Generalization ​

Author: Changliang Zhou, Canhong Yu, Shunyu Yao, Xi Lin, Zhenkun Wang, Yu Zhou, Qingfu Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.23413v3 Announce Type: replace Abstract: Multi-task neural routing solvers have emerged as a promising paradigm for their ability to solve multiple vehicle routing problems (VRPs) using a single model. However, existing neural solvers typically rely on predefined problem constraints or re...

📖 Read original article


173. TimePre: Bridging Accuracy, Efficiency, and Stability in Probabilistic Time-Series Forecasting ​

Author: Lingyu Jiang, Lingyu Xu, Peiran Li, Dengzhe Hou, Qianwen Ge, Dingyi Zhuang, Shuo Xing, Wenjing Chen, Xiangbo Gao, Ting-Hsuan Chen, Xueying Zhan, Xin Zhang, Ziming Zhang, Zhengzhong Tu, Michael Zielewski, Kazunori Yamada, Fangzhou Lin
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2511.18539v3 Announce Type: replace Abstract: We propose TimePre, a simple framework that unifies the efficiency of Multilayer Perceptron (MLP)-based models with the distributional flexibility of Multiple Choice Learning (MCL) for Probabilistic Time-Series Forecasting (PTSF). Stabilized Instan...

📖 Read original article


174. Delays in Spiking Neural Networks: A State Space Model Approach ​

Author: Sanja Karilanova, Subhrakanti Dey, Ay\c{c}a "Oz\c{c}elikkale
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.01906v3 Announce Type: replace Abstract: Spiking neural networks (SNNs) are biologically inspired, event-driven models suited for temporal data processing and energy-efficient neuromorphic computing. In SNNs, richer neuronal dynamic allows capturing more complex temporal dependencies, wit...

📖 Read original article


175. Auto-exploration for online reinforcement learning ​

Author: Caleb Ju, Guanghui Lan
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC

arXiv:2512.06244v4 Announce Type: replace Abstract: The exploration-exploitation dilemma in reinforcement learning (RL) is a fundamental challenge to efficient RL algorithms. Existing algorithms for finite state and action discounted RL problems address this by assuming sufficient exploration over b...

📖 Read original article


176. Clarity: The Flexibility-Interpretability Trade-Off in Sparsity-aware Concept Bottleneck Models ​

Author: Konstantinos P. Panousis, Diego Marcos
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.21944v3 Announce Type: replace Abstract: The widespread adoption of deep learning models in computer vision has intensified concerns about interpretability. Despite strong performance, these models are often treated as black boxes, with limited systematic investigation of their decision-m...

📖 Read original article


177. Putting a Face to Forgetting: Continual Learning meets Mechanistic Interpretability ​

Author: Sergi Masip, Gido M. van de Ven, Javier Ferrando, Tinne Tuytelaars
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.22012v3 Announce Type: replace Abstract: Catastrophic forgetting in continual learning is often measured at the performance or last-layer representation level, overlooking the underlying mechanisms. We introduce a mechanistic framework that offers a geometric interpretation of catastrophi...

📖 Read original article


178. CADET: Context-Conditioned Ads CTR Prediction With a Decoder-Only Transformer ​

Author: David Pardoe, Neil Daftary, Miro Furtado, Aditya Aiyer, Yu Wang, Liuqing Li, Tao Song, Lars Hertel, Young Jin Yun, Senthil Radhakrishnan, Zhiwei Wang, Tommy Li, Khai Tran, Ananth Nagarajan, Ali Naqvi, Yue Zhang, Renpeng Fang, Avi Romascanu, Arjun Kulothungun, Deepak Kumar, Praneeth Boda, Fedor Borisyuk, Ruoyan Wang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.11410v2 Announce Type: replace Abstract: Click-through rate (CTR) prediction is fundamental to online advertising systems. While Deep Learning Recommendation Models (DLRMs) with explicit feature interactions have long dominated this domain, recent advances in generative recommenders have ...

📖 Read original article


179. Order Matters in Retrosynthesis: Structure-aware Generation via Reaction-Center-Guided Discrete Flow Matching ​

Author: Chenguang Wang, Zihan Zhou, Lei Bai, Tianshu Yu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.13136v2 Announce Type: replace Abstract: Template-free retrosynthesis methods treat the task as black-box sequence generation, limiting learning efficiency, while semi-template approaches rely on rigid reaction libraries that constrain generalization. We address this gap with a key insigh...

📖 Read original article


180. Learning Representations from Incomplete EHR Data with Dual-Masked Autoencoding ​

Author: Xiao Xiang, David Restrepo, Hyewon Jeong, Yugang Jia, Leo Anthony Celi
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.15159v2 Announce Type: replace Abstract: Electronic health records (EHR) arrive masked. Clinicians order measurements selectively, and any patient table thus contains only a subset of the values that characterize the underlying physiological state. Prior masked modeling approaches on EHR ...

📖 Read original article


181. Risk-Averse Wasserstein Distributionally Robust Online Learning ​

Author: Guixian Chen, Salar Fattahi, Soroosh Shafiee
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2602.20403v2 Announce Type: replace Abstract: We study distributionally robust online learning, where a risk-averse learner updates decisions sequentially to guard against worst-case distributions drawn from a Wasserstein ambiguity set centered at past observations. While this paradigm is well...

📖 Read original article


182. Learning Disease-Sensitive Latent Interaction Graphs From Noisy Cardiac Flow Measurements ​

Author: Viraj Patel, Marko Grujic, Philipp Aigner, Theodor Abart, Marcus Granegger, Deblina Bhattacharjee, Katharine Fraser
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.23035v2 Announce Type: replace Abstract: Cardiac blood flow patterns contain rich information about disease severity and clinical interventions, yet current imaging and computational methods fail to capture underlying relational structures of coherent flow features. We propose a physics-i...

📖 Read original article


183. Exact and Asymptotically Complete Robust Verifications of Neural Networks via Ising Solvers ​

Author: Wenxin Li, Wenchao Liu, Weihao Li, Chuan Wang, Qi Gao, Yin Ma, Hai Wei, Kai Wen
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, physics.optics, quant-ph

arXiv:2603.00408v4 Announce Type: replace Abstract: We present an Ising-compatible framework for formal neural-network robustness verification under bounded input perturbations. For piecewise-linear activations, the Exact Logarithmic PWL Model (Log-PWL) provides an exact, sound, and complete formula...

📖 Read original article


184. Can Computational Reducibility Lead to Transferable Models for Graph Combinatorial Optimization? ​

Author: Semih Cant"urk, Thomas Sabourin, Frederik Wenkel, Michael Perlmutter, Guy Wolf
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.02462v2 Announce Type: replace Abstract: A key challenge in developing unified neural solvers for combinatorial optimization (CO) is the efficient generalization of models from a given set of tasks to new tasks unseen during initial training. To address this, we first establish a new GNN ...

📖 Read original article


185. Temporal Straightening for Latent Planning ​

Author: Ying Wang, Oumayma Bounou, Gaoyue Zhou, Randall Balestriero, Tim G. J. Rudner, Yann LeCun, Mengye Ren
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.12231v3 Announce Type: replace Abstract: Learning good representations is essential for latent planning with world models. While pretrained visual encoders produce strong semantic visual features, they are not tailored to planning and contain information irrelevant -- or even detrimental ...

📖 Read original article


186. Lost in Aggregation: On a Fundamental Expressivity Limit of Message-Passing Graph Neural Networks ​

Author: Eran Rosenbluth
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CC

arXiv:2603.14846v4 Announce Type: replace Abstract: We define an information-complexity property for aggregation functions, capturing a vast range of practical aggregations, and prove that any Message-Passing Graph Neural Network (MP-GNN) model with such aggregations induces only a polynomial number...

📖 Read original article


187. Lipschitz Dueling Bandits over Continuous Action Spaces ​

Author: Mudit Sharma, Shweta Jain, Vaneet Aggarwal, Ganesh Ghalme
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.IR, cs.MA

arXiv:2604.00523v2 Announce Type: replace Abstract: We study for the first time, stochastic dueling bandits over continuous action spaces with Lipschitz structure, where feedback is purely comparative. While dueling bandits and Lipschitz bandits have been studied separately, their combination has re...

📖 Read original article


188. BiScale-GTR: Fragment-Aware Graph Transformers for Multi-Scale Molecular Representation Learning ​

Author: Yi Yang, Ovidiu Daescu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.06336v2 Announce Type: replace Abstract: Fragment-level representations provide a natural way to capture recurring molecular substructures and reuse their learned representations across molecules. However, a shared fragment identity alone may not fully describe how a fragment is instantia...

📖 Read original article


189. Validated Synthetic Patient Generation for Small Longitudinal Cohorts: Coagulation Dynamics Across Pregnancy ​

Author: Jeffrey D. Varner, Maria Cristina Bravo, Carole McBride, Thomas Orfeo, Ira Bernstein
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM

arXiv:2604.07557v2 Announce Type: replace Abstract: Small longitudinal cohorts, common in maternal health, rare diseases, and early-phase trials, limit computational modeling because enrollment is slow and the data are too sparse to train reliable models. We present multiplicity-weighted Stochastic ...

📖 Read original article


190. Predictive Entropy as a Joint Screen for Error and Paraphrase Instability in Medical Vision-Language Models ​

Author: Binesh Sadanandan, Vahid Behzadan
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.08941v2 Announce Type: replace Abstract: Medical Vision-Language Models (VLMs) answering binary presence questions on chest radiographs can fail in two linked ways: they are confidently wrong, and they change answers when a clinically equivalent question is rephrased. In a binary answer h...

📖 Read original article


191. A Tale of Two Temperatures: Simple, Efficient, and Diverse Sampling from Diffusion Language Models ​

Author: Theo X. Olausson, Metod Jazbec, Xi Wang, Armando Solar-Lezama, Christian A. Naesseth, Stephan Mandt, Eric Nalisnick
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.09921v2 Announce Type: replace Abstract: Much work has been done on designing fast and accurate sampling for diffusion language models (dLLMs). However, these efforts have largely focused on the tradeoff between speed and quality of individual samples; how to additionally ensure diversity...

📖 Read original article


192. Correction and Corruption: A Two-Rate View of Error Flow in LLM Protocols ​

Author: Fernando Reitich
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.18245v3 Announce Type: replace Abstract: Large language models operate in protocols containing multiple calls, yet added calls are usually evaluated only by their net effect. That summary cannot distinguish correcting unsuccessful outputs from corrupting initially successful ones. We deve...

📖 Read original article


193. Scaling Self-Play with Self-Guidance ​

Author: Luke Bailey, Kaiyue Wen, Kefan Dong, Tatsunori Hashimoto, Tengyu Ma
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.20209v2 Announce Type: replace Abstract: LLM self-play algorithms are notable in that, in principle, nothing bounds their learning: a Conjecturer model creates problems for a Solver, and both improve together. However, in practice, existing LLM self-play methods do not scale well with lar...

📖 Read original article


194. From Local to Cluster: A Unified Framework for Causal Discovery with Latent Variables ​

Author: Zongyu Li
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.22416v3 Announce Type: replace Abstract: Latent variables pose a fundamental obstacle to both causal discovery and inference. Local approaches exploiting direct neighborhood relations provide little beyond immediate dependencies. Cluster-level methods, though capable of broader reasoning,...

📖 Read original article


195. Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models ​

Author: Cyril Shih-Huan Hsu, Wig Yuan-Cheng Cheng, Chrysa Papagianni
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.DC, cs.NI

arXiv:2604.26508v2 Announce Type: replace Abstract: Deploying Vision-Language Models (VLMs) on edge devices remains challenging due to their substantial computational and memory demands, which exceed the capabilities of resource-constrained embedded platforms. Conversely, fully offloading inference ...

📖 Read original article


196. Proteo-R1: Reasoning Foundation Models for De Novo Protein Design ​

Author: Fang Wu, Weihao Xuan, Heli Qi, Hanqun Cao, Heng-Jui Chang, Zeqi Zhou, Haokai Zhao, Ma Jian, Carl Ma, Yu-Chi Cheng, Kuan Pang, Xiangru Tang, Zehong Wang, Guanlue Li, Hanchen Wang, Kejun Ying, Pan Lu, Chiho Im, Seungju Han, Peng Xia, Tinson Xu, Yinxi Li, Deyao Zhu, Pheng-Ann Heng, Naoto Yokoya, Masashi Sugiyama, Li Erran Li, Jure Leskovec, Yejin Choi
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE

arXiv:2605.02937v2 Announce Type: replace Abstract: Deep learning in de novo protein design has achieved atomic-level fidelity. However, existing models remain largely non-deliberative: they directly synthesize molecular geometries without explicitly reasoning about which residues or interactions ar...

📖 Read original article


197. Same Targets, Different Computation: How Post-Training Divides Work Across Model Layers ​

Author: Yifan Zhou
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.07284v2 Announce Type: replace Abstract: A late-layer change learned during post-training may work on the base model's earlier state, or it may depend on earlier computation learned with it. We distinguish these cases with a four-cell diagnostic that crosses base or descendant upstream st...

📖 Read original article


198. Instance-Adaptive Online Multicalibration ​

Author: Zhiming Huang, Jamie Morgenstern, Aaron Roth, Claire Jie Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.09273v3 Announce Type: replace Abstract: We study online multicalibration beyond the worst-case. We give a single, efficient algorithm which dynamically interpolates between benign and worst-case sequences by adaptively refining a dyadic grid of prediction values. Its error is controlled ...

📖 Read original article


199. ConTact: Contact-First Antibody CDR Design via Explicit Interface Reasoning ​

Author: Mansoor Ahmed, Spencer VonBank, Nadeem Taj, Sujin Lee, Naila Jan, Murray Patterson
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.21600v3 Announce Type: replace Abstract: Computational antibody CDR design methods condition on antigen structure to generate binding loops. Yet, the existing architectures conflate two fundamentally distinct sub-problems: identifying which CDR positions will contact the antigen, and sele...

📖 Read original article


200. AgForce Enables Antigen-conditioned Generative Antibody Design ​

Author: Mansoor Ahmed, Murray Patterson
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.21610v2 Announce Type: replace Abstract: Antibody design methods condition on antigen structure to generate complementarity-determining regions (CDR), yet a systematic evaluation of baseline methods reveals that they largely ignore the antigen input. We identify three failure modes that e...

📖 Read original article


201. The Matching Principle: When Does a Training Penalty Cover Deployment Shift? ​

Author: Vishal Rajput
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2605.22800v3 Announce Type: replace Abstract: Ordinary training optimises the task loss and then stops. It never pays for internal representation energy: Jacobians can stay large in directions that never helped the label, so even small label-preserving noise throws the model off---a design gap...

📖 Read original article


202. Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robustness ​

Author: Manish Aryal, Faiyaz Azam, Agnivo Banerjee, Syed Mahir Ahamed, Sai Sidhanth Manoharan Jayanthi, Allegra Laro, Cl'ement Legentilhomme, Andrew Lin, Florian Lorkowski, Marina P'erez del Valle, Radman Rakhshandehroo, Patric Rommel, Emanuel Ruzak, Nathan Theng, Paul Yushin Rapoport
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.23146v3 Announce Type: replace Abstract: Classical reinforcement learning assumes the agent interacts with a fixed environment whose behavior does not depend on the agent's policy. This assumption breaks down in non-realizable settings where other actors might anticipate the agent's behav...

📖 Read original article


203. When Interpretability Is Unequally Distributed: Fairness in Hybrid Interpretable Models ​

Author: Ziba Jabbar Zare, Ulrich A"ivodji, Julien Ferry, Thibaut Vidal
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.28626v2 Announce Type: replace Abstract: Hybrid interpretable models combine a transparent component with a black-box model by assigning some examples to the former and deferring the rest to the latter. While this design enables flexible tradeoffs between accuracy and interpretability, it...

📖 Read original article


204. LVCG: Learning ECG Representations in the Latent Vectorcardiogram Space ​

Author: Bosong Huang, Panzhen Zhao, Zengxiang Li, Patricia Lee, Wei Jin, Alan Wee-Chung Liew, Ming Jin, Shirui Pan
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.31249v2 Announce Type: replace Abstract: Electrocardiography (ECG) is a cornerstone of cardiac assessment, making the learning of informative ECG representations fundamental to tasks ranging from disease diagnosis to clinical report generation. However, existing methods operate almost exc...

📖 Read original article


205. On Effectiveness and Efficiency of Agentic Tool-calling and RL Training ​

Author: Tong Liu, Cheng Qian, Matej Cief, Yuan He, Daniele Dan, Nikolaos Aletras, Gabriella Kazai
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.00135v2 Announce Type: replace Abstract: Tool-calling is a central component of modern large language model (LLM) agents, equipping them with skills beyond their parametric knowledge. This paper studies tool-calling along two complementary axes: effectiveness, i.e., how this capability is...

📖 Read original article


206. Model Multiplicity and Predictive Arbitrariness in Recidivism Risk Assessment ​

Author: Ashwin Singh, Carlos Castillo
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2606.02198v2 Announce Type: replace Abstract: Prediction tasks over individual futures, which are inherently noisy, often admit multiple similarly accurate models. When these models produce different predictions for the same individual, they raise concerns of arbitrariness in decision-making. ...

📖 Read original article


207. Where Flow Matching Leaks: Characterising Membership Signals Along the Interpolation Path ​

Author: Thomas Sesmat, Gabriel Meseguer-Brocal, Geoffroy Peeters
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SD

arXiv:2606.07271v3 Announce Type: replace Abstract: Understanding memorization in generative models remains challenging, with implications for copyright and privacy. Beyond verbatim reproduction, models can encode subtler traces of their training data that never surface in their outputs yet remain e...

📖 Read original article


208. Population-Aware Physics-Informed Neural Particle Flow for Robust Spacecraft Bayesian Navigation ​

Author: Batu Candan, Simone Servadio
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.10959v2 Announce Type: replace Abstract: Spacecraft navigation often requires Bayesian inference from sparse nonlinear measurements that produce curved, multimodal, or geometrically constrained posterior distributions. Physics-informed neural particle flow (PINPF) addresses such problems ...

📖 Read original article


209. Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation ​

Author: Amir El-Ghoussani, Michele De Vita, Ronald Naumann, Vasileios Belagiannis
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.11990v3 Announce Type: replace Abstract: Remaining Useful Life (RUL) prediction is essential for industrial predictive maintenance, yet many learning-based approaches rely on extensive feature engineering or large labeled datasets to train task-specific sequence models. In this work, we i...

📖 Read original article


210. Causal Variational Deep Embedding: A Family of Interventional Generators for Confounded Images ​

Author: Jingyuan Chen, Kangrui Ruan, Junzhe Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.21806v2 Announce Type: replace Abstract: Deep generative models reproduce the observational distribution of their training data, inheriting any spurious associations it contains. A common source is an unobserved confounder that shapes both an attribute the user wants to control at samplin...

📖 Read original article


211. KrishokChat: A Provenance-Traceable Multi-Task Bengali Agricultural Benchmark with Safety-Critical Chemical Advisory ​

Author: Khan Raiyan Ibne Reza, Sumaiya Tabassum Nimi, Omar Ibne Shahid
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.29243v2 Announce Type: replace Abstract: We introduce KrishokChat, an 85,979-instance Bengali agricultural benchmark built from 284 government publications, 13 institutions, and six regional dialects. The benchmark comprises four tracks: General Knowledge QA, Treatment QA, Safety Refusal ...

📖 Read original article


212. Foundations of Equivariant Deep Learning: Unifying Graph and Sheaf Neural Networks ​

Author: Yoshihiro Maruyama
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.03798v4 Announce Type: replace Abstract: Symmetry is everywhere in nature and society. Geometric deep learning builds architectures respecting group symmetries, whereas topological deep learning organizes computation through cells, incidence relations, and local-to-global structure. In th...

📖 Read original article


213. An Exact Instrument for State Usage in Selective State-Space Models, and the Input-Driven Migration It Reveals ​

Author: Raktim Bhattacharya
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11796v2 Announce Type: replace Abstract: Selective state-space models such as Mamba route information through a bank of first-order modes whose input coupling is set by a learned selection mechanism. We give an exact instrument for measuring how a trained model uses these modes. Because t...

📖 Read original article


214. Seq2Synth: Benchmarking Temporal Fidelity in Synthetic Sequential Tabular Data ​

Author: Kiwan Kwon, Kangmin Kim, Hojin Lee, Yeseong Jung, Hyeongwoo Kong, Vamsi K. Potluru, Saerom Park, Yongjae Lee
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.15606v2 Announce Type: replace Abstract: Synthetic sequential tabular data are increasingly used for privacy-preserving data sharing and data-driven research, but evaluating their fidelity remains difficult because temporal structure is easily lost under conventional tabular metrics. Exis...

📖 Read original article


215. Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions ​

Author: David R. Wessels, Farhad Ramezanghorbani, Alireza Moradzadeh, David W. Romero, Olivia Viessmann, Maksim Zhdanov, John St. John, Ken Janik, David M Knigge, Yucheng Tang, Erik J Bekkers, Saee Gopal Paliwal
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, stat.ML

arXiv:2607.19378v3 Announce Type: replace Abstract: Subquadratic alternatives to attention require compromises when applied to multi-dimensional data: standard convolutions lack global receptive fields and input dependency, while recurrent models require rasterizing data such as images, volumes, and...

📖 Read original article


216. An Insight on Evaluation Metrics Under the Imbalanced Case of Anomaly Detection ​

Author: Romain Hermary, Nesryne Mejri, Djamila Aouada
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.22286v2 Announce Type: replace Abstract: Anomaly detection is inherently characterised by severe class imbalance, making the interpretation of evaluation metrics challenging. Although metrics such as AUROC, AUPR, F1-score, and MCC are widely used, their values convey different meanings de...

📖 Read original article


217. Testing when adaptive data acquisition can replace fixed measurement plans ​

Author: Jia Bi, Samuel Pinilla, Chenyang Zhu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.27651v2 Announce Type: replace Abstract: Learned rules select samples for follow-up measurements in high-throughput experiments. Predicted value does not justify replacing a fixed plan. We introduce the opportunity-aware protocol for authorizing learned measurement rules (Opal), which lea...

📖 Read original article


218. Leak It: Per-Document Extraction Beyond Aggregate Membership Inference ​

Author: Victor Maricato
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CR

arXiv:2608.00144v2 Announce Type: replace Abstract: Membership inference (MIA) on language models is usually summarised by aggregate ROC-AUC, but such evaluations are confounded: model-free blind baselines can separate members from non-members using surface text alone. Building on probabilistic disc...

📖 Read original article


219. Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't ​

Author: Ravi Satya Durga Prasad Yenugula
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.02829v2 Announce Type: replace Abstract: Model families are typically trained size by size, each from scratch. Can apretrained large model instead be converted into a smaller sibling? Wecharacterize the 1.4B->410M conversion in the Pythia family end to end.Representations align strongly a...

📖 Read original article


220. Comparing SGLD and a fixed-noise Predictor-Corrector adaptation in canonical Joint Energy-Based Models on CIFAR-10 ​

Author: Dmytro Knopov
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.05025v2 Announce Type: replace Abstract: Joint Energy-Based Models (JEM) unify classification and generation within a single network and support out-of-distribution (OOD) detection. Canonical JEM training relies on stochastic gradient Langevin dynamics (SGLD); a theoretically motivated al...

📖 Read original article


221. Recent advances in weakly supervised learning: New supervision paradigms, assumption relaxations, and practical solutions ​

Author: Wei Wang, Gang Niu, Masashi Sugiyama
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.06896v2 Announce Type: replace Abstract: Deep learning has achieved great success in recent years thanks to the availability of high-quality, well-annotated training data. However, this requirement is often not met in real-world applications. Weakly supervised learning aims to train an ac...

📖 Read original article


222. Adaptive Supervised Anchoring for On-Policy Self-Distillation ​

Author: Meilin Yang (Renmin University of China, Beijing, China), Zixuan Ding (Renmin University of China, Beijing, China), Jianhao Nie (Renmin University of China, Beijing, China), Weite Zhang (Renmin University of China, Beijing, China), Yuxin Zhang (Renmin University of China, Beijing, China), Zhiming Shao (Renmin University of China, Beijing, China), Li Yu (Renmin University of China, Beijing, China), Zhe Fu (Renmin University of China, Beijing, China)
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07935v2 Announce Type: replace Abstract: On-policy self-distillation (OPSD) adapts a language model by distilling guidance from a frozen teacher on trajectories sampled from the student. Its effectiveness, however, depends critically on the quality of those trajectories. We show that when...

📖 Read original article


223. Robust Reputation-Driven Crowdsourced Federated Learning ​

Author: Mouhamed Amine Bouchiha, Gregory Blanc
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.DC

arXiv:2608.08574v2 Announce Type: replace Abstract: Crowdsourced Federated Learning (CrowdFL) extends traditional federated learning by enabling open and heterogeneous participation through a crowdsourcing paradigm. In this setting, reputation-driven incentive mechanisms are commonly employed to gui...

📖 Read original article


224. Measuring and Reducing WebGPU Dispatch Overhead for LLM Inference ​

Author: J\k{e}drzej Maczan
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.PF

arXiv:2608.08730v2 Announce Type: replace Abstract: Large Language Models are deployed to multiple types of environments, from internet browsers to edge devices, and WebGPU serves as a modern cross-platform standard. The engines for browser-based LLM inference have proliferated, yet the overhead of ...

📖 Read original article


225. Imaginative Generative AI: Crossing the Entropy Wall into Worlds Beyond Imitation ​

Author: Farzan Farnia, Hossein Goli, Amin Gohari
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.09385v2 Announce Type: replace Abstract: Generative AI models are primarily designed to imitate the data distribution, an objective that neither corrects diversity lost by a learned generator nor defines how generation should extend beyond the diversity of the data itself. We introduce Im...

📖 Read original article


226. Why Post-Norm Transformers Collapse: Attention Amplification and Gradient Repair Failure ​

Author: Xingjian Wang, Qingyu Han, Xiaodong Luo, Yin Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09417v2 Announce Type: replace Abstract: Deep decoder-only Transformers often replace the original Post-Norm architecture with Pre-Norm variants because Post-Norm training is highly sensitive to warmup and learning rate under conventional initialization schemes. Although prior work has id...

📖 Read original article


227. Bayesian Symbolic Regression with Entropic Reinforcement Learning ​

Author: Oussama Boussif, Mohammed Mahfoud, Younesse Kaddar, Moksh Jain, Sida Li, Damiano Fornasiere, Xiaoyin Chen, Yoshua Bengio, Esmeralda S. Whitammer
Published: 8/12/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09617v2 Announce Type: replace Abstract: Symbolic regression is the problem of finding an algebraic expression describing a stochastic dependence of a target variable on a set of inputs. Unlike forms of regression that fit parameters assuming a fixed model structure, symbolic regression i...

📖 Read original article


228. Emergent Neural Network Mechanisms for Generalization to Objects in Novel Orientations ​

Author: Avi Cooper, Xavier Boix, Daniel Harari, Spandan Madan, Hanspeter Pfister, Tomotake Sasaki, Pawan Sinha
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, q-bio.NC, stat.ML

arXiv:2109.13445v3 Announce Type: replace-cross Abstract: The capability of Deep Neural Networks (DNNs) to recognize objects in orientations outside the distribution of the training data is not well understood. We present evidence that DNNs are capable of generalizing to objects in novel orientation...

📖 Read original article


229. Representation and Invariance in Reinforcement Learning ​

Author: Samuel Alexander, Arthur Paul Pedersen
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.GT, cs.LG

arXiv:2112.07752v5 Announce Type: replace-cross Abstract: Researchers have formalized reinforcement learning (RL) in different ways. If an agent in one RL framework is to run within another RL framework's environments, the agent must first be converted, or mapped, into that other framework. In this ...

📖 Read original article


230. A variational Bayes approach to inference for low-dimensional parameters in high-dimensional linear regression ​

Author: Isma"el Castillo, Alice L'Huillier, Kolyan Ray, Luke Travis
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2406.12659v3 Announce Type: replace-cross Abstract: We propose a scalable variational Bayes method for statistical inference for a single or pre-specified low-dimensional subset of the coordinates of a high-dimensional parameter in sparse linear regression. Our approach relies on assigning a m...

📖 Read original article


231. Regression and Classification with Single-Qubit Quantum Neural Networks ​

Author: Leandro C. Souza, Bruno C. Guingo, Gilson Giraldi, Renato Portugal
Published: 8/12/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG

arXiv:2412.09486v2 Announce Type: replace-cross Abstract: The literature reflects a mutually beneficial relationship between machine learning and quantum computing, where progress in one field frequently drives improvements in the other. Motivated by the rich connection between these areas, we use a...

📖 Read original article


232. KKL Observer Synthesis for Nonlinear Systems via Physics-Informed Learning ​

Author: M. Umar B. Niazi, John Cao, Matthieu Barreau, Karl Henrik Johansson
Published: 8/12/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2501.11655v3 Announce Type: replace-cross Abstract: This paper proposes a novel learning approach for designing Kazantzis-Kravaris or nonlinear Luenberger (KKL) observers for autonomous nonlinear systems. The design of a KKL observer involves finding an injective map that transforms the system...

📖 Read original article


233. Robust and Secure Code Watermarking for Large Language Models via ML/Crypto Codesign ​

Author: Ruisi Zhang, Neusha Javidnia, Nojan Sheybani, Farinaz Koushanfar
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG

arXiv:2502.02068v3 Announce Type: replace-cross Abstract: This paper introduces RoSeMary, the first-of-its-kind ML/Crypto codesign watermarking framework that regulates LLM-generated code to avoid intellectual property rights violations and inappropriate misuse in software development. High-quality ...

📖 Read original article


234. Bayesian Federated Cause-of-Death Classification and Quantification Under Distribution Shift ​

Author: Yu Zhu, Jason Teng, Zehang Richard Li
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.AP

arXiv:2505.02257v2 Announce Type: replace-cross Abstract: In regions lacking medically certified causes of death, verbal autopsy (VA) is a widely used tool to ascertain the cause of death through interviews with caregivers. Data collected by VAs are often analyzed using probabilistic algorithms. The...

📖 Read original article


235. Local Fr\'echet functional regression in manifolds from time-correlated bivariate curve data ​

Author: M. D. Ruiz-Medina, A. Torres-Signes
Published: 8/12/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.ML, stat.TH

arXiv:2505.05168v4 Announce Type: replace-cross Abstract: Under mild conditions, a least-squares local linear Fr'echet curve predictor is derived for a response and a regressor evaluated in a separable Hilbert space. The conditions that allow the implementation of the local linear Fr'echet functio...

📖 Read original article


236. Generalized Linear Markov Decision Process ​

Author: Sinian Zhang, Kaicheng Zhang, Ziping Xu, Zongqi Xia, Jue Hou, Tianxi Cai, Doudou Zhou
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2506.00818v2 Announce Type: replace-cross Abstract: Offline reinforcement learning for longitudinal studies often faces two linked challenges: rewards may be binary or bounded, and reward observations may be available only for a subset of trajectories or time points even when the corresponding...

📖 Read original article


237. FARCLUSS: Fuzzy Adaptive Rebalancing and Contrastive Uncertainty Learning for Semi-Supervised Semantic Segmentation ​

Author: Ebenezer Tarubinga, Jenifer Kalafatovich, Seong-Whan Lee
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV

arXiv:2506.11142v3 Announce Type: replace-cross Abstract: Semi-supervised semantic segmentation (SSSS) faces persistent challenges in effectively leveraging unlabeled data, such as ineffective utilization of pseudo-labels, exacerbation of class imbalance biases, and neglect of prediction uncertainty...

📖 Read original article


238. Smooth Flow Matching for Synthesizing Functional Data ​

Author: Jianbin Tan, Anru R. Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2508.13831v4 Announce Type: replace-cross Abstract: Functional data, i.e., random functions observed over a continuous domain, are increasingly available in areas such as biomedical research, health informatics, and epidemiology. However, effective statistical analysis for functional data is o...

📖 Read original article


239. Behavioral Inference at Scale: The Fundamental Asymmetry Between Motivations and Belief Systems ​

Author: Jason Starace, Terence Soule
Published: 8/12/2026, 4:00:00 AM
Categories: cs.MA, cs.LG

arXiv:2509.05624v3 Announce Type: replace-cross Abstract: How much information about an agent's underlying values can be recovered from its observable behavior? This question matters for any approach that infers agent properties from action sequences, yet remains empirically open at scale. We addres...

📖 Read original article


240. Learning in ImaginationLand: Omnidirectional Policies through 3D Generative Models (OP-Gen) ​

Author: Yifei Ren, Edward Johns
Published: 8/12/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2509.06191v2 Announce Type: replace-cross Abstract: Recent 3D generative models, which are capable of generating full object shapes from just a few images, now open up new opportunities in robotics. In this work, we show that 3D generative models can be used to augment a dataset from a single ...

📖 Read original article


241. Faster Results from a Smarter Schedule: Reframing Collegiate Cross Country through Analysis of the National Running Club Database ​

Author: Jonathan A. Karr Jr, Ryan M. Fryer, Nitesh V. Chawla
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG

arXiv:2509.10600v5 Announce Type: replace-cross Abstract: Collegiate cross country teams often build their season schedules on intuition rather than evidence, partly because large-scale performance datasets were not publicly accessible prior to the National Running Club Database (NRCD). We analyze t...

📖 Read original article


242. Diffusion-Based Impedance Learning for Contact-Rich Manipulation Tasks ​

Author: Noah Geiger, Tamim Asfour, Neville Hogan, Johannes Lachner
Published: 8/12/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2509.19696v4 Announce Type: replace-cross Abstract: Learning-based methods excel at robot motion generation but remain limited in contact-rich physical interaction. Impedance control provides stable and safe contact behavior but requires task-specific tuning of stiffness and damping parameters...

📖 Read original article


243. On The Statistical Limits of Self-Improving Agents ​

Author: Charles L. Wang, Keir Dorchen, Peter Jin
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2510.04399v3 Announce Type: replace-cross Abstract: We develop a learning-theoretic framework for analyzing self-improving agents by decomposing self-modification into five axes. Within this framework, we prove a sharp boundary: under standard i.i.d. assumptions, distribution-free PAC learnabi...

📖 Read original article


244. HyWA: Architecture-Preserving Personalized Voice Activity Detection for Full-Duplex Voice Assistants ​

Author: Hamed Jafarzadeh Asl, Amin Edraki, Mahsa Ghazvini Nejad, Masoud Asgharian, Mohammadreza Sadeghi, Yuanhao Yu, Vahid Partovi Nia
Published: 8/12/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.LG, cs.SD

arXiv:2510.12947v3 Announce Type: replace-cross Abstract: Voice activity detection (VAD) serves as an early gate in voice-assistant pipelines for smart devices. Because conventional VADs respond to speech from any speaker, nearby conversations and residual assistant playback lead to unwanted trigger...

📖 Read original article


245. Concept Labels Are Not Enough: Rethinking Concept Bottleneck Models through Representation Integrity ​

Author: Gaoxiang Huang, Songning Lai, Yutao Yue
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2510.15770v4 Announce Type: replace-cross Abstract: Although deep neural networks achieve strong predictive performance, their internal reasoning often remains difficult to inspect and control. Concept Bottleneck Models (CBMs) address this opacity by factoring predictions through human-underst...

📖 Read original article


246. Gradient Boosted Mixed Models: Flexible Estimation of Mean and Variance Components for Clustered Data ​

Author: Mitchell L. Prevett, Francis K. C. Hui, Zhi Yang Tho, A. H. Welsh, Anton H. Westveld
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO, stat.ME

arXiv:2511.00217v3 Announce Type: replace-cross Abstract: We introduce a novel way to combine gradient boosting with mixed effects models, whereby the mean and variance components are learned jointly as functions of covariates via likelihood-based gradients. Gradient Boosted Mixed Models (GBMixed) e...

📖 Read original article


247. A Streaming Sparse Cholesky Method for Derivative-Informed Gaussian Process Surrogates Within Digital Twin Applications ​

Author: Shridhar Vashishtha, Krishna Prasath Logakannan, Jacob Hochhalter, Shandian Zhe, Robert M. Kirby
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.CE, cs.LG

arXiv:2511.00366v3 Announce Type: replace-cross Abstract: Digital twins are developed to model the behavior of a specific physical asset (or twin), and they can consist of high-fidelity physics-based models or surrogates. A highly accurate surrogate is often preferred over multi-physics models as th...

📖 Read original article


248. VDC-Agent: When Video Detailed Captioners Evolve Themselves via Agentic Self-Reflection ​

Author: Qiang Wang, Xinyuan Gao, Yuhang He, Jizhou Han, Jiangyang Li, SongLin Dong, Zhiheng Ma, Yihong Gong
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MM

arXiv:2511.19436v2 Announce Type: replace-cross Abstract: Existing Video Detailed Captioning (VDC) methods predominantly rely on costly human annotations or distillation from powerful proprietary models, creating a dependency on external supervision. In this paper, we propose VDC-Agent, an autonomou...

📖 Read original article


249. On the Condition Number Dependency in Bilevel Optimization ​

Author: Lesi Chen, Kaiyi Ji, Jingzhao Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: math.OC, cs.AI, cs.LG

arXiv:2511.22331v4 Announce Type: replace-cross Abstract: Bilevel optimization minimizes an objective function, defined by an upper-level problem whose feasible region is the solution of a lower-level problem. We study the oracle complexity of finding an $\epsilon$-stationary point with first-order ...

📖 Read original article


250. Automated Data Enrichment using Confidence-Aware Fine-Grained Debate among Open-Source LLMs for Mental Health and Online Safety ​

Author: Junyu Mao, Anthony Hills, Talia Tseriotou, Maria Liakata, Aya Shamir, Dan Sayda, Dana Atzil-Slonim, Natalie Djohari, Pamela Ugwudike, Mahesan Niranjan, Stuart E. Middleton
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2512.06227v3 Announce Type: replace-cross Abstract: Real-world indicators play an important role in many Natural Language Processing (NLP) applications, such as life events for mental health analysis and risky behaviours for online safety, yet labelling such information is often costly and/or ...

📖 Read original article


251. On Solomonoff Induction in Large Language Models and the Limits of Self-Improving: The Singularity Is Not Near Without Symbolic Model Synthesis ​

Author: Hector Zenil
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, math.IT

arXiv:2601.05280v3 Announce Type: replace-cross Abstract: On the one hand, the question of whether large language models (LLMs) are Solomonoff induction estimators has become an explicit question at the intersection of Algorithmic Information Theory (AIT) and Machine Learning (ML) of great interest....

📖 Read original article


252. Nonlinear multi-study sparse factor analysis ​

Author: Gemma E. Moran, Anandi Krishnan
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2601.18128v2 Announce Type: replace-cross Abstract: High-dimensional data often exhibit variation that can be captured by lower-dimensional factors. For high-dimensional data from multiple studies, one goal is to understand which underlying factors are common to all studies, and which factors ...

📖 Read original article


253. Enhancing Automated Essay Scoring With Three Techniques: Two-Stage Fine-Tuning, Score Alignment, and Self-Training ​

Author: Hongseok Choi, Serynn Kim, Wencke Liermann, Jin Seong, Jin-Xia Huang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2602.01747v2 Announce Type: replace-cross Abstract: Automated Essay Scoring (AES) plays a crucial role in education by providing scalable and efficient assessment tools. However, in real-world settings, the extreme scarcity of labeled data severely limits the development and practical adoption...

📖 Read original article


254. Bandwidth-Efficient Multi-Agent Communication through Information Bottleneck and Vector Quantization ​

Author: Ahmad Farooq, Kamran Iqbal
Published: 8/12/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.IT, cs.LG, cs.MA, math.IT

arXiv:2602.02035v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning systems deployed in real-world robotics applications face severe communication constraints that significantly impact coordination effectiveness. We present a framework that combines information bottleneck th...

📖 Read original article


255. LLMs Encode Their Failures: Predicting Success from Pre-Generation Activations ​

Author: William Lugoloobi, Thomas Foster, William Bankes, Chris Russell
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2602.09924v4 Announce Type: replace-cross Abstract: Running LLMs with extended reasoning on every problem is expensive, but determining which inputs actually require additional compute remains challenging. We investigate whether their own likelihood of success is recoverable from their interna...

📖 Read original article


256. Efficient Uncoupled Learning Dynamics with $\tilde{O}\!\left(T^{-1/4}\right)$ Last-Iterate Convergence in Bilinear Saddle-Point Problems over Convex Sets under Bandit Feedback ​

Author: Arnab Maiti, Claire Jie Zhang, Kevin Jamieson, Jamie Heather Morgenstern, Ioannis Panageas, Lillian J. Ratliff
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.GT, cs.LG

arXiv:2602.21436v2 Announce Type: replace-cross Abstract: In this paper, we study last-iterate convergence of learning algorithms in bilinear saddle-point problems, a preferable notion of convergence that captures the day-to-day behavior of learning dynamics. We focus on the challenging setting wher...

📖 Read original article


257. MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games ​

Author: Jacob Eisenstein, Fantine Huot, Adam Fisch, Jonathan Berant, Mirella Lapata
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2602.24188v2 Announce Type: replace-cross Abstract: We present a scalable and verifiable methodology for evaluating language models in multi-turn interactions, using a suite of collaborative games that require effective communication about private information. This enables an interactive scali...

📖 Read original article


258. Dual Space Preconditioning for Gradient Descent in the Overparameterized Regime ​

Author: Reza Ghane, Danil Akhtiamov, Babak Hassibi
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2603.10485v3 Announce Type: replace-cross Abstract: In this work, we study the convergence properties of the Dual Space Preconditioned Gradient Descent, encompassing optimizers such as Normalized Gradient Descent and Gradient Clipping. We consider preconditioners of the form $\nabla K$, where ...

📖 Read original article


259. UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference ​

Author: Lang Zhou, Shuxuan Li, Zhuohao Li, Shi Liu, Zhilin Zhao, Wei-Shi Zheng
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2603.18446v2 Announce Type: replace-cross Abstract: Long-context inference remains challenging for large language models due to attention dilution and out-of-distribution degradation. Context selection mitigates this limitation by attending to a subset of key-value cache entries, yet most meth...

📖 Read original article


260. Discrete Coefficients and Open Invariant Covers in the Homology of Ample Groupoids ​

Author: Luciano Melodia
Published: 8/12/2026, 4:00:00 AM
Categories: math.AT, cs.LG, math.KT, math.OA

arXiv:2603.20861v2 Announce Type: replace-cross Abstract: The homology of an ample groupoid is computed from the complex of compactly supported continuous functions on the nerve. Two hypotheses routinely imposed on this complex behave in opposite ways. We show that the comparison map from the integr...

📖 Read original article


261. UniScale: Synergistic Entire Space Data and Model Scaling for Search Ranking ​

Author: Liren Yu, Caiyuan Li, Feiyi Dong, Tao Zhang, Zhixuan Zhang, Dan Ou, Haihong Tang, Bo Zheng
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2603.24226v4 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have inspired a surge of scaling research in industrial search, advertising, and recommendation systems. However, existing approaches focus mainly on architectural improvements, overlooking the ...

📖 Read original article


262. Does Explanation Correctness Matter? Linking Computational XAI Evaluation to Human Understanding ​

Author: Gregor Baer, Chao Zhang, Isel Grau, Pieter Van Gorp
Published: 8/12/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.LG

arXiv:2603.25251v2 Announce Type: replace-cross Abstract: Explainable AI (XAI) methods are commonly evaluated using functional correctness metrics, sometimes termed faithfulness or fidelity, which estimate how closely an explanation reflects the model's reasoning. Higher correctness is assumed to pr...

📖 Read original article


263. Reinforcement Learning-based Semi-supervised Knowledge Distillation with LLM-as-a-Judge ​

Author: Yiyang Shen, Lifu Tu, Weiran Wang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2604.02621v2 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) substantially improves the reasoning capabilities of language models, but most existing RL fine-tuning approaches rely entirely on ground-truth verifiable rewards and thus labeled datasets with verifiable answers. ...

📖 Read original article


264. RankFormer: A Propose-then-Select Transformer for Multi-Agent Multimodal Trajectory Prediction ​

Author: Diyi Liu, Zihan Niu, Tu Xu, Xingchen Zhang, Lishan Sun
Published: 8/12/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2604.07126v3 Announce Type: replace-cross Abstract: Predicting traffic agent trajectories plays an important role in autonomous driving, traffic operations, transportation safety analysis, etc. Although many deep learning algorithms are devised to predict future agent trajectories, the traject...

📖 Read original article


265. Loop, Think, & Generalize: Implicit Reasoning in Recurrent-Depth Transformers ​

Author: Harsh Kohli, Srinivasan Parthasarathy, Huan Sun, Yuekun Yao
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.07822v2 Announce Type: replace-cross Abstract: We study implicit reasoning, i.e. the ability to combine knowledge or rules within a single forward pass. While transformer-based large language models store substantial factual knowledge and rules, they often fail to compose this knowledge f...

📖 Read original article


266. Inverse Design of Inorganic Compounds with Generative AI ​

Author: Hannes Kneiding, Luc'ia Mor'an-Gonz'alez, Nishamol Kuriakose, Ainara Nova, David Balcells
Published: 8/12/2026, 4:00:00 AM
Categories: physics.chem-ph, cond-mat.mtrl-sci, cs.LG

arXiv:2604.11827v2 Announce Type: replace-cross Abstract: Machine learning is revolutionizing chemistry. Beyond the value of predictive models accelerating virtual screening, generative AI aims at enabling inverse design, reversing the compound-to-property prediction paradigm into property-to-compou...

📖 Read original article


267. Simpler Logarithmic Approximation Algorithms for the Optimal Decision Tree and Adaptive Set Cover ​

Author: Micha{\l} Szyfelbein
Published: 8/12/2026, 4:00:00 AM
Categories: cs.DS, cs.IR, cs.LG

arXiv:2604.12036v3 Announce Type: replace-cross Abstract: We study a well-known task of constructing a decision tree identifying an unknown hypothesis from a given ground set of hypotheses under both the average- and worst-case cost. The Optimal Decision Tree problem has been extensively studied in ...

📖 Read original article


268. Null-Space Flow Matching for MIMO Channel Estimation in Latency-Constrained Systems ​

Author: Junjie Zhao, Guangming Liang, Xiaonan Liu, Dongzhu Liu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, eess.SP, math.IT

arXiv:2604.22005v2 Announce Type: replace-cross Abstract: Accurate yet low-latency channel state information (CSI) acquisition is essential for multiple-input multiple-output (MIMO) communication systems. While advanced deep generative models, such as score-based and diffusion models, enable high-fi...

📖 Read original article


269. The Exact Replica Threshold for Nonlinear Moments of Quantum States ​

Author: Shuai Zeng
Published: 8/12/2026, 4:00:00 AM
Categories: quant-ph, cs.CC, cs.IT, cs.LG, math.IT, physics.comp-ph

arXiv:2604.22627v2 Announce Type: replace-cross Abstract: Joint measurements on multiple copies of a quantum state provide access to nonlinear observables such as $\operatorname{tr}(\rho^t)$, but whether replica number marks a sharp information-theoretic resource boundary has remained unclear. For e...

📖 Read original article


270. Spherical Flows for Sampling Categorical Data ​

Author: Jannis Chemseddine, Gregor Kornhardt, Gabriele Steidl
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.CL, cs.LG

arXiv:2605.05629v4 Announce Type: replace-cross Abstract: We study the problem of learning generative models for discrete sequences in a continuous embedding space. Whereas prior approaches typically operate in Euclidean space or on the probability simplex, we instead work on the sphere $\mathbb S^{...

📖 Read original article


271. Choosing a JPEG Decoder for PyTorch DataLoaders: Workload-Specific Throughput on Four CPUs ​

Author: Vladimir Iglovikov, Dmitry Kosarevsky
Published: 8/12/2026, 4:00:00 AM
Categories: cs.PF, cs.LG

arXiv:2605.08731v3 Announce Type: replace-cross Abstract: A JPEG decoder benchmark can combine worker counts, CPUs, and datasets in one large result matrix. We simplify that comparison by fixing a PyTorch DataLoader at eight workers and asking one question: how much faster is each decoder than Pillo...

📖 Read original article


272. On the global convergence of gradient flow for wide shallow models beyond homogeneous nonlinearities ​

Author: Romain Petit, Clarice Poon, Gabriel Peyr'e
Published: 8/12/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2605.10775v2 Announce Type: replace-cross Abstract: A surprising phenomenon in the training of neural networks is the ability of gradient descent to find global minimizers of the training loss despite its non-convexity. Following earlier work, we investigate this behavior for wide shallow mode...

📖 Read original article


273. Grounded Post-Training with Hard Examples for Reducing Hallucination in Multimodal Large Language Models ​

Author: Qinwu Xu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.DB, cs.LG

arXiv:2605.16411v3 Announce Type: replace-cross Abstract: Hallucination remains a fundamental challenge in vision-language models (VLMs), where autoregressive generation may produce linguistically plausible yet physically inconsistent or visually ungrounded responses due to likelihood maximization u...

📖 Read original article


274. Why Do Safety Guardrails Degrade Across Languages? ​

Author: Max Zhang, Ameen Patel, Sang T. Truong, Sanmi Koyejo
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2605.17173v2 Announce Type: replace-cross Abstract: Large language models exhibit safety degradation in non-English languages. Standard evaluation relies on Jailbreak Success Rate (JSR), which confounds several safety-driving factors into one, obscuring the specific cause(s) of safety failure....

📖 Read original article


275. HoloQ-VLA: Uniform W4A4 Quantization of Vision-Language-Action Models ​

Author: Xinyu Wang, Mingze Li, Sicheng Lyu, Dongxiu Liu, Kaicheng Yang, Ziyu Zhao, Yufei Cui, Xiao-Wen Chang, Peng Lu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2605.28803v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models unify perception, reasoning, and control in a single policy, but their multi-billion-parameter backbones and diffusion-based action heads make on-device deployment prohibitively expensive. Low-bit post-trai...

📖 Read original article


276. Taming the Loss Landscape of PINNs with Noisy Feynman-Kac Supervision: Operator Preconditioning and Non-Asymptotic Error Bounds ​

Author: Nathanael Tepakbong, Hanyu Hu, Chengyu Liu, Xiang Zhou
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.NA, math.OC, math.ST, stat.TH

arXiv:2606.00643v2 Announce Type: replace-cross Abstract: Physics-Informed Neural Networks (PINNs) often train slowly or fail to converge on challenging partial differential equations (PDEs), a behavior recently linked to severely ill-conditioned loss landscapes inherited from the underlying differe...

📖 Read original article


277. Moxia: A Trust-First Neuro-Symbolic Execution Architecture for Self-Explaining Mathematical Reasoning ​

Author: Alessio Bruno
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2606.00671v3 Announce Type: replace-cross Abstract: We present Moxia (formerly AXIOM), a trust-first neuro-symbolic architecture for self-explaining mathematical reasoning over natural-language input. Its language model is strictly a canonicalizer: it rewrites informal problem text into a narr...

📖 Read original article


278. A case study of evaluating AI agents on a neuroscience data-to-discovery pipeline ​

Author: Kai A. Horstmann, Ethan Lin, Alice A. Robie, Jennifer J. Sun, Kristin Branson
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG

arXiv:2606.07718v2 Announce Type: replace-cross Abstract: Agentic AI offers a promising path to automating software development bottlenecks in scientific research pipelines, particularly for stages that take domain experts days to months to build and where correctness and robustness matter more than...

📖 Read original article


Author: Yan Dai, Maryam Farboodi, Negin Golrezaei, Sepehr Shahshahani
Published: 8/12/2026, 4:00:00 AM
Categories: econ.TH, cs.AI, cs.GT, cs.LG, stat.ML

arXiv:2606.12260v3 Announce Type: replace-cross Abstract: How can we design a market of human-generated content for use in training AI models that both enables technological progress and preserves individual incentives for high-quality content creation? Existing approaches take polar positions: a "f...

📖 Read original article


280. Masked Neural Detection for Run-Length-Limited Channel Coding in Molecular Communication ​

Author: Melih \c{S}ahin, Ozgur B. Akan
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT

arXiv:2606.12489v2 Announce Type: replace-cross Abstract: Molecular communication (MC) suffers from severe diffusion memory because molecules released for one symbol may arrive during later symbol intervals. Neural sequence detectors, especially sliding bidirectional recurrent neural networks (SBRNN...

📖 Read original article


281. The Truth Stays in the Family: Enhancing Contextual Grounding via Inherited Truthful Heads in Model Lineages ​

Author: Miso Choi, Seonga Choi, Mincheol Kwon, Woosung Joung, Jinkyu Kim, Jungbeom Lee
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2606.15821v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have produced many specialized multimodal LLMs (MLLMs) that share common foundational LLMs, forming distinct model lineages. It remains unclear whether a fundamental behavioral link exists betwe...

📖 Read original article


282. Forecasting With LLMs: Improved Generalization Through Feature Steering ​

Author: Humzah Merchant, Bradford Levy
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2606.27199v2 Announce Type: replace-cross Abstract: Successful forecasting involves identifying patterns between historical and future states of the world which generalize to future observations. We apply LLMs to a variety of forecasting tasks and inspect their internal states using sparse aut...

📖 Read original article


283. The Calibrated Deepfake Trust Score (CDTS): Competence-Coupled Trust Degradation Across Deepfake Detectors ​

Author: Md Anas Biswas
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CR, cs.CV, cs.LG

arXiv:2606.29484v2 Announce Type: replace-cross Abstract: In moderation, provenance, and verification pipelines a deepfake detector's output probability is read as a degree of trust, so its calibration matters as much as raw accuracy. We reframe deepfake detection as a calibrated, self-auditing trus...

📖 Read original article


284. Coachable agents for interactive gameplay ​

Author: Roberto Capobianco (Sony AI, Zurich, Switzerland), Harm van Seijen (Sony AI, North America, various locations), Nolan D. Bard (Sony AI, North America, various locations), Neil Burch (Sony AI, North America, various locations), Fatima Davelouis (Sony AI, North America, various locations), Josh Davidson (Sony AI, North America, various locations), Alisa Devlic (Sony AI, Zurich, Switzerland), Yunshu Du (Sony AI, North America, various locations), Ishan Durugkar (Sony AI, North America, various locations), Siddhant Gangapurwala (Sony AI, North America, various locations), Daniel Hernandez (Sony AI, North America, various locations), G. Zacharias Holland (Sony AI, North America, various locations), Sahil Jain (Sony AI, North America, various locations), Kenta Kawamoto (Sony AI, Tokyo, Japan), Raksha Kumaraswamy (Sony AI, North America, various locations), Patrick MacAlpine (Sony AI, North America, various locations), Dustin R. Morrill (Sony AI, North America, various locations), Declan Oller (Sony AI, North America, various locations), Francesco Riccio (Sony AI, Zurich, Switzerland), Akanksha Saran (Sony AI, North America, various locations), Craig Sherstan (Sony AI, Tokyo, Japan), Kaushik Subramanian (Sony AI, Zurich, Switzerland), Thomas J. Walsh (Sony AI, North America, various locations), Samuel Barrett (Sony AI, North America, various locations), Kizza N. Frisbee (Sony AI, North America, various locations), Mady Govil (Sony AI, North America, various locations), Johannes G"unther (Sony AI, North America, various locations), Varun R. Kompella (Sony AI, North America, various locations), James A. MacGlashan (Sony AI, North America, various locations), Maxwell Svetlik (Sony AI, North America, various locations), Michael D. Thomure (Sony AI, North America, various locations), Jaden B. Travnik (Sony AI, North America, various locations), Kevin Waugh (Sony AI, North America, various locations), Elahe Aghapour (Sony AI, North America, various locations), Florian Fuchs (Sony AI, Zurich, Switzerland), Andreanne Lemay (Sony AI, North America, various locations), Shruti Mishra (Sony AI, Zurich, Switzerland), Takuma Seno (Sony AI, Tokyo, Japan), Peter Stone (Sony AI, North America, various locations), Michael Spranger (Sony AI, Tokyo, Japan), Peter R. Wurman (Sony AI, North America, various locations)
Published: 8/12/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.00642v2 Announce Type: replace-cross Abstract: Reinforcement learning has proven to be a valuable tool in the creation of advanced AI and robotic systems, contributing to everything from game playing to robotics to foundation models. Through trial-and-error, these AI systems typically lea...

📖 Read original article


285. What You See Is What You Get: Observation-Aligned Supervision for Chart-to-Code Generation ​

Author: Tianhao Niu, Qingfu Zhu, Wanxiang Che
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.04726v3 Announce Type: replace-cross Abstract: Chart-to-code generation is commonly trained with supervised fine-tuning on reference plotting scripts, implicitly treating the gold code as a fully observable target. We argue that this assumption is often invalid: many chart programs contai...

📖 Read original article


286. TSCoNet: A Two-Stage Copula CNN-LSTM for Uncertainty-Aware Spatio-Temporal Forecasting ​

Author: Jongwook Kim, Jong-Min Kim
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.10410v2 Announce Type: replace-cross Abstract: Reliable forecasting of several interrelated environmental variables - such as regional precipitation and temperature, or other correlated geophysical fields - across many locations calls for accurate predictions accompanied by trustworthy st...

📖 Read original article


287. Adaptive Filtering of the KV Cache: Diagnosing and Correcting Structural-Role Bias in LLM Inference ​

Author: Soumil Mandal
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.13205v2 Announce Type: replace-cross Abstract: Attention-based KV cache eviction (H2O and its descendants) compresses the memory-constrained state of a long-context model by ranking tokens on accumulated attention mass, treated here as signal energy, and keeping the heaviest. On schema-de...

📖 Read original article


288. Position: The Inevitable Transition to Machine Learning in Quantum Chemistry ​

Author: Karen Sargsyan, Chao-Ping Hsu
Published: 8/12/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, quant-ph

arXiv:2607.18281v3 Announce Type: replace-cross Abstract: Finding exact solutions to the quantum many-body problem is computationally intractable (QMA-hard). Traditional approximations for electrons in an atom or molecule -- density functional theory and wavefunction methods -- have been indispensab...

📖 Read original article


289. SearchArt: Training Long-Horizon Search Agent with Scalable Synthetic and Verified Task ​

Author: Lang Mei, Xiaohan Yu, Chong Chen, Liyan Liu, Xiangnan Chen, Jinchao Ma, Chao Feng, Li Huang, Siyu Mo, Sichen Kang, Yunkun Xu, Zhihan Yang, Zhujun Xue, Jingren Zhang, Qing He, Yingdi Huang, Hao Jiang, Ziao Ma, Zewei Pan, Minhao Sun, Zhuo Tao, Jinzhao Xiao, Gangtao Xin, Huanyao Zhang, Wenjian Zhang, Jiangshan Zhang, Guojie Zhu, Fangzhou Zou, Jiaxin Mao, Wentao Zhang
Published: 8/12/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.24850v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have enabled search agents to autonomously tackle complex tasks across extended search and reasoning horizons. However, training effective search agents remains challenging due to the lack of sc...

📖 Read original article


290. AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents ​

Author: Ruoyu Wang, Heng Zhao, Renjie Wu, Mengnan Zhao, Zhixuan Chu, Wanyu Lin, Tianhang Zheng
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG

arXiv:2607.26998v3 Announce Type: replace-cross Abstract: Large language model (LLM) agents automate penetration testing through an observation-action loop, selecting actions based on observations returned by tools. This dependence allows defenders to inject deceptive observations that can mislead t...

📖 Read original article


291. Policy-Masked Private Experts: Auditable and Reversible Capability Access Control in Sparse MoE Models ​

Author: Zhuoheng Huang, Mukesh Singh
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.06690v2 Announce Type: replace-cross Abstract: Most language-model access controls regulate behavior while leaving the same computation available to every request. We study a different systems question: can trusted authorization determine which newly trained parameters are reachable by th...

📖 Read original article


292. Physics-Informed Condition Monitoring of SiC Power Modules ​

Author: Mattia Scarpa, Evgeny Kusmenko, Francesco Toso, Mattia Bruschetta, Ruggero Carli, Simon Achatz
Published: 8/12/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2608.08363v2 Announce Type: replace-cross Abstract: Silicon carbide (SiC) power modules are increasingly deployed in automotive traction inverters, where condition monitoring is essential to prevent in-service failures. Despite extensive qualification under AQG 324, no consolidated approach ex...

📖 Read original article


293. Failure-Mechanism Transferability of Cumulative-Damage Features for Health State Estimation of SiC Power Modules ​

Author: Mattia Scarpa, Evgeny Kusmenko, Francesco Toso, Mattia Bruschetta, Ruggero Carli, Simon Achatz
Published: 8/12/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2608.08365v2 Announce Type: replace-cross Abstract: Data-driven health-state estimators for SiC (Silica-Carbide) power modules typically report their performance on a single accelerated-aging campaign, and how that performance transfers to a different failure mechanism is rarely tested. We ben...

📖 Read original article


294. Population-Level Generative Modeling for Ranking Data ​

Author: Zhaoyang Shi
Published: 8/12/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, math.ST, stat.ML, stat.TH

arXiv:2608.08422v2 Announce Type: replace-cross Abstract: Ranking data arise in scientific and machine learning applications, including recommendation systems, information retrieval, voting, marketing, and AI preference ranking from human feedback. Existing statistical work has primarily focused on ...

📖 Read original article


295. OpenVisTool: An Open Recipe for Synthesizing Instructive Visual Tool-Use Trajectories ​

Author: Changhao Xiang, Shilin Zhang, Zheng Ma, Kanzhi Cheng, Ruize Ma, Yi Feng, Jianbing Zhang, Zhi Wang, Zhen Wu, Xinyu Dai, Lewei Lu
Published: 8/12/2026, 4:00:00 AM
Categories: cs.CL, cs.CV, cs.LG

arXiv:2608.08557v2 Announce Type: replace-cross Abstract: Visual tool use has emerged as a fundamental capability for multimodal agents to actively acquire evidence beyond a fixed image encoding. The prevailing recipe learns this capability from teacher-generated trajectories filtered for answer cor...

📖 Read original article