Skip to content

arXiv cs.LG - 2026-08-13 ​

261 items collected.


1. FarSky: Task-Aware Latent-Space Coupling for Generative Intra-Hour Solar Forecasting ​

Author: Yann Fabel, Bijan Nouri, Milon Miah, Niklas Blum, Luis F. Zarzalejo, Julia Kowalski, Robert Pitz-Paal
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph, stat.AP

arXiv:2608.11254v1 Announce Type: new Abstract: Accurate solar irradiance forecasting is essential for the reliable integration of photovoltaic power into modern electricity grids. All-sky imagers (ASI) provide high-resolution observations of clouds, making them well suited for intra-hour forecastin...

📖 Read original article


2. Why AI Detection Fails for Academic Integrity ​

Author: Jonathan A. Karr Jr, Grigorii Khvatskii, Ting Hua, Nitesh V. Chawla
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2608.11256v1 Announce Type: new Abstract: Institutions use commercial AI detectors for academic integrity, yet detectors cannot distinguish AI editing from full LLM drafts and may treat both as misconduct. In a controlled study of published English abstracts (four domains; 2013 to 2015 vs. 202...

📖 Read original article


3. Basin: Efficient and Extensible Numerical Optimization in Rust ​

Author: Johan Larsson
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.11279v1 Announce Type: new Abstract: Basin is a numerical optimization library for the Rust programming language. Numerical optimization is the task of finding the inputs that minimize a function, and it is a fundamental element across the sciences: fitting a model to data, calibrating a ...

📖 Read original article


4. Federated Learning for Distributed CNC Tool Wear Prediction ​

Author: Afsana Khan, Morris Stallmann, Marcin Pietrasik, Charis Kouzinopoulos, Anna Wilbik
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11281v1 Announce Type: new Abstract: Tool wear prediction is an important task in CNC machining, where accurate monitoring of tool condition supports product quality and process reliability. Machine learning methods have shown potential for this task, but their use in industrial environme...

📖 Read original article


5. Terminal Symmetry as a Decision Resource: Statewise Refinement for Anytime Verified Construction ​

Author: Yi Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11318v1 Announce Type: new Abstract: Many sequential construction tasks exhibit exact symmetry at completion while their execution remains directed and history-dependent. We develop a decision-resource view of terminal symmetry: process evidence supplies directionality, terminal correspon...

📖 Read original article


6. Contextual Quality-Diversity Evolutionary Reinforcement Learning for HVAC Control in Tropical Commercial Buildings ​

Author: Tran Le Vu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE, math.OC

arXiv:2608.11324v1 Announce Type: new Abstract: This paper proposes a contextual quality-diversity evolutionary reinforcement-learning controller, CQD-ERL, for the supervisory control of a tropical, water-cooled chiller plant and its associated air side. Rather than converging to a single scalarised...

📖 Read original article


7. Long-Horizon Forecasting of Complete Financial Statements with Forma ​

Author: Travis L. Johnson, Jiannan Jiang, Soumyabrata Chaudhuri, Yihao Chen, Lauren Falvey, Donal O'Cofaigh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP

arXiv:2608.11327v1 Announce Type: new Abstract: Specialist training beats generalist scale when forecasting financial statements. To our knowledge, no prior work jointly forecasts complete financial statements beyond one year, yet in a discounted-cash-flow valuation most firm value sits past that wi...

📖 Read original article


8. Weightless Fine-Tuning: Personalizing LLMs via Logit-Space Transport ​

Author: Bohan Zhang, Anqi Ni, Yixin Wang, Paramveer S. Dhillon
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.11342v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is a standard approach for adapting LLMs to a target distribution, but in settings such as personalization, where each author requires separate weight access, optimization, storage, and retraining, its costs become prohibit...

📖 Read original article


9. Dynamics Models for Offline Hyperparameter Selection in Real-World RL ​

Author: Jordan Coblin, Han Wang, Martha White, Adam White
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11349v1 Announce Type: new Abstract: A key obstacle to deploying reinforcement learning in real-world systems is hyperparameter selection, particularly when simulators are unavailable and online experimentation is costly. Prior work has proposed calibration models trained on offline data ...

📖 Read original article


10. Market-Information-Aware Gated-LoRA of Foundation Models for Transferable Day-Ahead Electricity Price Forecasting ​

Author: Hang Fan, Wei Wei, Shengwei Mei
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11359v1 Announce Type: new Abstract: Electricity price forecasting is crucial for market participants but remains difficult because prices are volatile, market-specific, and closely tied to anticipated system conditions. Existing supervised methods depend largely on market-specific histor...

📖 Read original article


11. Lifecycle-Optimal Tokenization: Vocabulary Size as a Deployment-Regime-Dependent Infrastructure Parameter ​

Author: Rima Mittal, Ankit Gubrani, Satyanarayana Kakollu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.PF

arXiv:2608.11361v1 Announce Type: new Abstract: Tokenizer vocabulary size is a foundational design choice in large language model (LLM) infrastructure, yet it is typically fixed at training time based on convention rather than deployment analysis. We show that the cost-optimal vocabulary is not a co...

📖 Read original article


12. PAIR: Pairwise-Aware Inclusion Reweighting for Adaptive Rollout Allocation in RLVR ​

Author: Pixel Nomand, Elena Voss, Marcus Hale, Sofia Reyes
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11368v2 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) spends most of its compute generating groups of long reasoning trajectories. Recent allocators reduce this cost by assigning budgets to prompts, rollouts, or tokens according to a pointwise notion o...

📖 Read original article


13. Towards an approach to multivariate outlier detection for District Heating System data ​

Author: Rajko Turudija, Du\v{s}an Stojiljkovi'c, Milan Zdravkovi'c, Marko Ignjatovi'c
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2608.11375v1 Announce Type: new Abstract: In this paper, we test different methods for multivariate detection of outliers in the data of transmitted heat energy in the selected substation of local District Heating System, by also considering outside ambient temperature, namely Z-score (univari...

📖 Read original article


14. Reoptimization Algorithms for Contextual Bandits with Knapsack Constraints ​

Author: Zhen Xu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.11383v1 Announce Type: new Abstract: We study new algorithms for Contextual Bandits with Knapsack. In these problems, there are finitely many types of customers, products, and resources. Each product is made from a fixed combination of resources, and resources have finite capacity. A deci...

📖 Read original article


15. Mechanism Design for Generative Engines: From Exploitation toward Win-Win Outcomes ​

Author: Chen Xu, Zitian Guo, Chenyan Xiong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11390v1 Announce Type: new Abstract: Generative engines are reshaping the web ecosystem by making citations a key mechanism for allocating attention, attribution, and downstream value. This creates a strategic tension: content providers are incentivized to optimize for model citation, whi...

📖 Read original article


16. Unmasking Toxic Mimicry in Medical Offline Reinforcement Learning for ICU Sepsis Management via Counterfactual Clinical Audits ​

Author: Hangqi Ren, Junyi Liao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2608.11410v1 Announce Type: new Abstract: Offline reinforcement learning (RL) offers considerable promise for optimizing ICU treatment decisions, yet standard evaluation metrics Mean Squared Error (MSE) and Fitted Q-Evaluation (FQE) assess only behavioral imitation and cannot detect Toxic Mimi...

📖 Read original article


17. Diffusion-Based Data-Driven Assortment Optimization ​

Author: Junyi Liao, Xiaohui Jiang, Zhengwei Tong, Ethan X. Fang, Vahid Tarokh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11419v1 Announce Type: new Abstract: Assortment optimization is a fundamental problem in revenue management, typically addressed using parametric choice models such as the multinomial logit (MNL) and its variants. While these models enable tractable formulations, their performance is sens...

📖 Read original article


18. Analysis of Federated Aggregation under Model Poisoning and Backdoor Attacks: A Reconstructed Cross-Dataset and Cross-Architecture Benchmark ​

Author: Soumya Mazumdar, Vineet Kumar Rakesh, Tapas Samanta
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.CV

arXiv:2608.11423v1 Announce Type: new Abstract: Robust comparisons of federated aggregation methods require joint consideration of predictive performance, threat definitions, metric semantics, and execution provenance. A 500-cell seed-1 evaluation matrix was reconstructed across five aggregation met...

📖 Read original article


19. Click2Poly: A VLM for vector mapping buildings and walls ​

Author: Nicolas Girard, Jawher Ben Abdallah, Arno Gobbin, Liuyun Duan, Sacha Lepretre
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.11424v1 Announce Type: new Abstract: Accurate vector mapping of buildings and walls is critical for geospatial applications but remains a labor-intensive process. While recent deep learning methods have improved automatic extraction, in order to meet cartographic standards they always req...

📖 Read original article


20. Three Tokens Force Exponential Feature Rank in Nonnegative Kernel Attention ​

Author: Vicente Opazo
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11427v1 Announce Type: new Abstract: Full attention exposes every token pair, whereas kernel attention compresses a sequence into a fixed-dimensional sketch. We show that this distinction becomes exponential at the first context length containing two competing candidates. On Min-IP over B...

📖 Read original article


21. AutoGrable: What Is a Good Graph for a Table? ​

Author: Tamara Cucumides, Floris Geerts
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11431v1 Announce Type: new Abstract: Graph learning presupposes a graph, and tables and relational databases do not come with one. Applying a GNN to them requires deciding which entities become nodes, which of them to connect, and through which relations---a decision made by hand, by sche...

📖 Read original article


22. Variational Parameter Calibration with Physics-Aware Latent-Space Surrogates ​

Author: Qiyao Zhou, Xujia Zhu, Pierre Joli, Yu Cong, Sibo Cheng
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph, physics.data-an, physics.flu-dyn

arXiv:2608.11435v1 Announce Type: new Abstract: Forward and inverse modeling of parametric dynamical systems requires surrogate models that are not only accurate for state prediction, but also informative for parameter calibration. However, a systematic end-to-end differentiable formulation for coup...

📖 Read original article


23. XGBoost "is all you need": the case of forecasting transmitted heat energy in District Heating Systems ​

Author: Milan Zdravkovi'c
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2608.11446v1 Announce Type: new Abstract: This paper presents a comparative study of two distinct approaches, XGBoost and Long-Short Term Memory (LSTM), for forecasting transmitted heat energy in District Heating Systems (DHS). The objective is to explore scenarios in which conventional ML alg...

📖 Read original article


24. PAC-Bayes Beyond Parameter Space: Behavioral Equivalence, Z-Information, and Exact Complexity Decomposition ​

Author: Vasant G. Honavar, Satish Kumar Keshri, Neil Ashtekar, Zehao Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11465v1 Announce Type: new Abstract: PAC-Bayes theory provides generalization guarantees by controlling the Kullback--Leibler (KL) divergence between posterior and prior distributions over a chosen hypothesis representation. However, predictive risk depends only on the predictive behavior...

📖 Read original article


25. Dual-Primal Graph VAEs for Noisy Label Aggregation ​

Author: Patrick Stinson, Nikolaus Kriegeskorte
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11473v1 Announce Type: new Abstract: Inferring the ground-truth from noisy crowdsourced labels is an important theoretical and practical problem. Neural network-based methods offer an alternative to classical Bayesian models which require specifying a family of generative models used for ...

📖 Read original article


26. Convergence Guarantees of Gradient Descent for Neural Networks via Generalized Lipschitz Smoothness ​

Author: Siqiao Mu, Diego Klabjan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11479v1 Announce Type: new Abstract: We establish convergence guarantees of gradient descent for general feedforward neural networks of arbitrary width or depth, with no special requirements on the initialization or dataset. We only assume that the activation functions are Lipschitz smoot...

📖 Read original article


27. Defending against Model Extraction for GNNs with Model Reprogramming ​

Author: Yan Wen, Zhenyi Wang, Heng Huang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.IR

arXiv:2608.11495v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) serve as the backbone for high-stakes applications in Machine-Learning-as-a-Service (MLaaS). Still, their black-box deployment exposes them to Model Extraction (ME) attacks, in which adversaries steal intellectual property ...

📖 Read original article


28. HyperFix: Combinatorial Nonlinear Correction for Task Vector Merging ​

Author: Hyo Seo Kim, Ren Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11499v1 Announce Type: new Abstract: Task vectors enable model merging without joint retraining. In practice, the subset of task vectors to be merged may vary, but many existing methods use scalar tuning for a particular subset, requiring repeated tuning across subsets and restricting tas...

📖 Read original article


29. RelShap: Relationally Consistent Shapley Explanations ​

Author: Seungeun Lee, Joao Fonseca, Julia Stoyanovich
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11508v1 Announce Type: new Abstract: Machine learning pipelines commonly flatten relational data into single-table representations, discarding structural constraints. Widely used Shapley value-based feature attributions then rely on feature independence, evaluating the model on combinatio...

📖 Read original article


30. Let it Cook: Learning to Wait in Sequential Decision Making ​

Author: Christopher Watson, Arjun Krishna, Dinesh Jayaraman, Rajeev Alur
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11511v1 Announce Type: new Abstract: In sequential decision making, an agent typically observes its environment and acts at every timestep. However, such active participation may not always be necessary; tasks such as brewing coffee include periods that are served equally well by letting ...

📖 Read original article


31. FLARE++: Low-rank attention with dynamic attention routing ​

Author: Vedant Puri, Yongjie Jessica Zhang, Levent Burak Kara
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11519v1 Announce Type: new Abstract: Full self-attention is a strong token mixer for PDE surrogates on irregular domains, but its quadratic cost limits its use on high-resolution problems. Efficient latent-attention models such as the Fast Low-rank Attention Routing Engine (FLARE) avoid t...

📖 Read original article


32. Hierarchical Federated Transfer Learning in Digital Twin-Based Vehicular Networks ​

Author: Qasim Zia, Saide Zhu, Haoxin Wang, Zafar Iqbal, Yingshu Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11532v1 Announce Type: new Abstract: In recent research on the Digital Twin-based Vehicular Ad hoc Network(DT-VANET), Federated Learning (FL) has shown its ability to provide data privacy. However, Federated learning struggles to adequately train a global model when confronted with data h...

📖 Read original article


33. Robust Ambiguity Detection (RAD) From Model- and Feature-Space Consistency ​

Author: Manya Singh, Mark T. Keane, Arjun Pakrashi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11541v1 Announce Type: new Abstract: Machine learning models should be robust, in the sense of remaining predictively consistent under permissible variations. A model's predictions should ideally remain unchanged when it is replaced by a functionally equivalent one, or when its inputs are...

📖 Read original article


34. Certifying What Helps Customer-Return Timing: A Screen-and-Confirm Test for Conditioning Signals, and Why Decay Is Nearly Enough ​

Author: Sang Su Lee, Vineeth Loganathan, Shishir Dash, Vijay Raghavan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.AP

arXiv:2608.11555v1 Announce Type: new Abstract: Practitioners enrich customer-return models with ever more signals (lifetime value, category, recency/frequency, calendar, geography), and the temporal-point-process (TPP) literature follows suit with covariate- and external-covariate-conditioned inten...

📖 Read original article


35. When Offline Evaluation Misleads: A Diagnostic Protocol for Reward and Policy Selection in Delayed-Feedback Contextual Bandits ​

Author: Sang Su Lee, Vineeth Loganathan, Shishir Dash, Vijay Raghavan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11560v1 Announce Type: new Abstract: Personalizing marketing messages with contextual multi-armed bandits (CMABs) drives real business value, yet the objective that ultimately matters - a downstream conversion - is observed only weeks later, too late to drive online learning. Teams theref...

📖 Read original article


36. Sparse and robust geometric twin support vector machine via asymmetric RoBoSS loss function ​

Author: Kai Qi, Xinji Huang, Hongchun Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11567v1 Announce Type: new Abstract: In real-world scenarios, the training data usually contains redundant features, label noise and feature noise, which provide severe challenges for the efficiency of machine learning methods. Since standard support vector machine (SVM) adopts $l_2$-norm...

📖 Read original article


37. RECAST: A Machine-Learning Framework for Correction and Super-Resolution of Coarse-Grid PDE Solvers ​

Author: Maryam Reza, Farbod Faraji
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph

arXiv:2608.11572v1 Announce Type: new Abstract: Coarse-grid numerical solvers can substantially reduce the computational cost of time-dependent PDE simulation, but under-resolution often degrades both the trajectory and the spatial fidelity of the solution. We introduce RECAST (Recurrent Error Corre...

📖 Read original article


38. Dion3: Full-Stack Orthogonal Updates ​

Author: Noah Amsel, Jack Zhang, Kwangjun Ahn, Ali Naeimi, Austin Feng, Berlin Chen, Tri Dao, John Langford
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11612v1 Announce Type: new Abstract: The Muon optimizer incurs a significant overhead cost due to its cubic-time Newton-Schulz orthogonalization step. When weights are sharded, communication overhead compounds this computational cost, eroding the benefits of Muon in many settings. We pres...

📖 Read original article


39. A Local Sinkhorn Framework for Conditional Distribution Reconstruction of Multidimensional Random Fields ​

Author: Mingtao Xia, Qijing Shen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11613v1 Announce Type: new Abstract: In this paper, we propose a local Sinkhorn divergence framework for conditional distribution reconstruction of multidimensional random fields. By utilizing the debiased Sinkhorn divergence, our proposed approach develops a differentiable and computatio...

📖 Read original article


40. FM-LLM: A frequency-enhanced mixture-of-experts framework for adapting LLMs to time series forecasting ​

Author: Rentao Gu, Yihang Ding, Junjie Li, Yi Ding, Weijing Sang, Xiaoli Huo, Xin Qin, Yuefeng Ji
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NI, eess.SP

arXiv:2608.11623v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have spurred cross-modal solutions for time-series forecasting. However, existing methods rely heavily on textual prompts for modality alignment-introducing nontrivial computational overhead and failing t...

📖 Read original article


41. Transferable Above-Ground Biomass (AGB) Estimation Model from Multi-Sensor Data with Sparse Field Calibration ​

Author: Pann Thinzar Seint, Bryan Atwood, Subas Chhatkuli
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.11638v1 Announce Type: new Abstract: Spatially continuous quantification of forest above-ground biomass (AGB) is what makes carbon accounting credible and mitigation strategies actionable. While field inventories provide high localized accuracy, they are spatially sparse; conversely, spac...

📖 Read original article


42. Towards a Formal Definition of Agent Memory: Basis, Span, Optimality, and the Sequential Memory Problem ​

Author: Hongyao Tang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11654v1 Announce Type: new Abstract: Despite the wide deployment of memory in large-model agents, there is no unified formal account of what a memory is or when it is optimal. This paper takes a first step toward this account. The central idea is that memory is a basis, knowledge is its s...

📖 Read original article


43. Continuous-Latent Predictive Modeling with Semantic Alignment for EEG-Language Foundation Models ​

Author: Myeong-Ju Cho, Hye-Bin Shin, Seo-Hyun Lee, Seong-Whan Lee
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11656v1 Announce Type: new Abstract: Recent advances in EEG foundation models have demonstrated the potential of large-scale pretraining to enable generalizable neural decoding across subjects, recording environments, and datasets. However, dominant pretraining paradigms face key challeng...

📖 Read original article


44. Is Per-Agent Policy Composition Safe? Rethinking Successor-Feature Transfer in Cooperative Multi-Agent Reinforcement Learning ​

Author: Zijian Zhao, Sen Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA

arXiv:2608.11658v1 Announce Type: new Abstract: Many reinforcement learning systems, from fleet management to traffic signal control, must serve an objective that changes dynamically after deployment, and retraining a policy for each new objective is prohibitively expensive. For a single agent, this...

📖 Read original article


45. Low-Interaction-Rank Learning: Unifying Multiplicative Dual-Encoder Heads ​

Author: Zijian Zhao, Sen Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11661v1 Announce Type: new Abstract: A multiplicative dual-encoder network computes a real-valued output for a pair of inputs as the inner product of their separate encodings. This architecture has been developed independently in operator learning, bipartite matching, contrastive vision-l...

📖 Read original article


46. Rubric Dropout: A Simple Way to Mitigate Reward Hacking in Rubric-as-Reward RL ​

Author: Minglai Yang, Xinyu Guo, Utkarsh Tyagi, Mian Zhang, Razvan Dumitru, Sunjie Hou, Yunzhong He, Daniel Yue Zhang, Ying Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.11669v1 Announce Type: new Abstract: Reinforcement learning against rubrics, lists of criteria graded by an LLM judge, has become a standard way to post-train language models on tasks with no deterministic answer. The rubric, however, is a fixed proxy for quality, never a complete descrip...

📖 Read original article


47. GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs ​

Author: Kai Yang, Jingwei Xu, Wanyu Wang, Kai-Yuan Guo, Zhenbo Yu, Yi Wang, Yu Qiao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11674v1 Announce Type: new Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they frequently suffer from training instabilities, cross-task capability degradation, and response-length inflation. Although prior work has characterize...

📖 Read original article


48. FunnelCausalNet: Funnel-aware Joint Conversion-Revenue Uplift for Multi-tier Coupon Allocation ​

Author: Yu Zhang (AMap Alibaba Group, Beijing, China), Zhihan Wang (AMap Alibaba Group, Beijing, China), Guanlin Chen (AMap Alibaba Group, Beijing, China), Min Jiang (AMap Alibaba Group, Beijing, China), Shuai Li (AMap Alibaba Group, Beijing, China)
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2608.11675v1 Announce Type: new Abstract: Coupon campaigns seek to lift both conversion and revenue, but gross merchandise value (GMV) follows a deterministic funnel from conversion to conditional order value and is zero-inflated and heavy-tailed. We propose FunnelCausalNet, an uplift estimato...

📖 Read original article


49. Drift and Dependence: Layer-wise Information-Theoretic Bounds for Replay-Based Continual Learning ​

Author: Tieliang Gong, Zhongbo Zhang, Wen Wen, Yong-Jin Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.11690v1 Announce Type: new Abstract: Continual learning must absorb new tasks without erasing old ones, and replay---mixing a small buffer of past examples into current training---is among the most effective remedies for catastrophic forgetting. Yet its generalization behavior is shaped b...

📖 Read original article


50. LEMUR: Latent Entropy-aware Multimodal Unlearning via Visual-anchored Reasoning Redirection ​

Author: Xinhao Zhong, Yuxia Qiao, Junhao Li, Hao Fang, Yi Sun, Bin Chen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.11691v1 Announce Type: new Abstract: Reinforcement-learning (RL) post-training equips multimodal large reasoning models (MLRMs) with exploratory chains of thought (CoT), substantially improving visual reasoning. However, we find that this capability introduces a distinct privacy vulnerabi...

📖 Read original article


51. REOPD: Reliability-Adaptive Reward Extrapolation for On-Policy Distillation ​

Author: Yang Sun, Lichao Ma, Houyuan Qin, Yuxin Liu, Hanyang Lu, Yao Zhu, Pinlong Cai, Guohang Yan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11698v2 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own trajectories under dense token-level supervision from a teacher. Reward-extrapolation methods such as ExOPD amplify the teacher-reference log-likelihood ratio to move beyond direct imitation, but...

📖 Read original article


52. Consolidator: Learning Persistent Routed Memory Across Context Boundaries ​

Author: Sungwoo Goo, Hwi-yeol Yun, Sangkeun Jung
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11701v1 Announce Type: new Abstract: Copying short-term memory (STM) into a slower store can preserve state across a context boundary, but persistence alone does not ensure that the retained state influences subsequent memory access. We test this distinction in a Phasor Memory Network (PM...

📖 Read original article


53. Robust and Efficient Noisy-Label Time-Series Classification via Dynamic Time Warping Based Granular Ball Computing ​

Author: Ziqiang Li, Yun Liu, Gouhei Tanaka
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11704v1 Announce Type: new Abstract: Dynamic Time Warping (DTW)-based Nearest-Neighbor (NN) classifiers are effective for time-series classification but are vulnerable to mislabeled training samples and require numerous DTW computations during inference. We propose DTW-based Granular Ball...

📖 Read original article


54. High-dimensional Multi-objective Bayesian Optimization with Learned Variable Interactions ​

Author: Hongyan Wang, Jiayu Huang, Haotian Zheng, Xin Gao, Chi Ding, Ying Liu, Xia Wang, Qing Xu, Keqiang Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11713v1 Announce Type: new Abstract: Multi-objective Bayesian optimization (MOBO) is effective in identifying the Pareto fronts for expensive black-box problems. However, most current MOBO approaches are limited to low-dimensional decision space due to its exponential sampling complexity....

📖 Read original article


55. Chain-of-Thought Shows the Path to a Tree: Realizing Branching Complexity ​

Author: Debanjan Dutta, Anish Chakrabarty, Swagatam Das
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11716v1 Announce Type: new Abstract: Chain of Thought (CoT) lifts the expressive ceiling of bounded-depth Transformers, with characterizations tying the number of CoT steps to circuit complexity classes. What remains largely missing are concrete instantiations with explicit, depth-bounded...

📖 Read original article


56. Epiplexity Guided Data Selection and Generation for Out-of-Distribution Generalization ​

Author: Ellen Su, Andres Potapczynski, Shikai Qiu, Edward Hughes, Andrew Gordon Wilson
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.11746v1 Announce Type: new Abstract: Modern systems are increasingly expected to transfer across tasks not specified during training. What data facilitates generalization in these new, unanticipated settings? One hypothesis is that data with more structural information could contain share...

📖 Read original article


57. MOON: Multi-Objective OrthoNormalized Updates for Multitask Learning ​

Author: Shiji Zhou, Kunlin Lyu, Lei Zhang, Ruodong Wang, Yifan Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.11749v1 Announce Type: new Abstract: Multi-objective optimization (MOO) has demonstrated significant success in multi-task learning by mitigating task conflicts through gradient manipulation. However, most existing methods flatten model parameters into vectors and perform gradient manipul...

📖 Read original article


58. TradingMoE: Routing the Right Experts in Evolving Markets ​

Author: Chang Zhou, Xingtong Yu, Minbin Huang, Zhennan Wu, Yuan Fang, Hong Cheng, Xinming Zhang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11785v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong potential for financial analysis and trading, but direct trading remains challenging because the predictive capabilities required can vary across assets, decision fields, and market conditions. Existing LL...

📖 Read original article


59. High-Order Liquid Evidence Encoding for Gradual GNSS Spoofing Detection in Autonomous Driving ​

Author: Muhammad Ayub Sabir, Junbiao Pang, Fatima Ashraf
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11790v1 Announce Type: new Abstract: Accurate Global Navigation Satellite System (GNSS)-based localization is essential for safe and reliable autonomous driving. However, spoofing attacks can manipulate vehicle position estimates. Continuous and subtle attacks are particularly difficult t...

📖 Read original article


60. Orientation, not magnitude: the causal structure of task-vector interference in merged language models ​

Author: Chencheng Zhu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11797v1 Announce Type: new Abstract: Model merging by task arithmetic works until it doesn't, and the field diagnoses why with magnitudes: layerwise representation bias, deviations from cross-task linearity, parameter overlap. Tracking the exact layerwise cross-term of merged LLMs through...

📖 Read original article


61. JAPE: Joint Anomaly Prediction and Intrinsic Explanation in Multivariate Time Series ​

Author: Yian Wei, Yuanyuan Yao, Lu Chen, Xiangmin Zhou, Tianyi Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11801v1 Announce Type: new Abstract: Multivariate time-series anomaly prediction aims to identify whether and when anomalies will occur over a future horizon from historical observations. Existing methods primarily characterize anomalies as deviations in future numerical values, which may...

📖 Read original article


62. Learning with Bilevel-Minimax Optimization for Efficient and Reliable Transfer Attacks ​

Author: Yaohua Liu, Yifan Guo, Jiaxin Gao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.11815v1 Announce Type: new Abstract: Transfer-based adversarial attacks craft adversarial examples using surrogate models to mislead black-box victim models. Beyond perturbation generation, transferability is fundamentally governed by the coupling of initialization, surrogate adaptation, ...

📖 Read original article


63. Towards Understanding On-Policy Distillation through the Lens of Test-Time Scaling ​

Author: Xinmu Ge, Zizhuo Zhang, Yu Huang, Jianing Zhu, Lin Yuan, Wanli Gu, Weichang Wu, Weiran Huang, Xiaolu Zhang, Bo Han, Jun Zhou, Jiangchao Yao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.11829v1 Announce Type: new Abstract: On-policy distillation (OPD) has emerged as a promising post-training technique for enhancing LLM reasoning. It is commonly believed to enable the student model to distill knowledge from a stronger teacher model, thereby expanding capabilities beyond t...

📖 Read original article


64. Kernel Methods for Learning Operators with Multiple Inputs and Outputs ​

Author: Adrien Weihs, Chunyang Liao, Jingmin Sun, Hayden Schaeffer
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH

arXiv:2608.11831v1 Announce Type: new Abstract: Learning mappings between infinite-dimensional objects is a central challenge in scientific machine learning. We introduce a general kernel-based encoder-decoder framework for operator learning that separates observation, representation, learning, and ...

📖 Read original article


65. Air Quality Station Simulation via LSTM and Attention-Based Modelling ​

Author: Alexander Kostadinov, Petar O. Hristov, Dessislava Petrova-Antonova
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11839v1 Announce Type: new Abstract: Poor air quality in urban areas is driven by a complex chain of processes and presents a significant public health concern. To better understand and control the mechanisms that determine air quality, cities deploy networks of measurement stations, and ...

📖 Read original article


66. Small-Scale Experiments: Are We There Yet? ​

Author: Nicholas Lourie, Kyunghyun Cho, Karen Ullrich, Sanae Lotfi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11859v1 Announce Type: new Abstract: Scaling laws promised cost-effective experiments; six years later, they have yet to fully deliver. Instead, researchers have found them unreliable at small scales (starting at 4M parameters) and concluded that sizable models cannot be avoided. We show ...

📖 Read original article


67. Forward and Inverse Virtual Metrology for Phototransistor Gain: A Hierarchical, Uncertainty-Aware Approach for Small Production Datasets ​

Author: Mahshid Amirabgir, Lorenza Ferrario, Paolo Conci, Mahdieh Amirabgir, Giancarlo Orengo
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.SY, eess.SY

arXiv:2608.11868v1 Announce Type: new Abstract: The customization, optimization and stabilization of the process flow of a silicon bipolar phototransistor commits months of cleanroom time before a finished device can be measured, so a model that predicts device gain from process parameters before a ...

📖 Read original article


68. DCM Bandits: Multiplayer Information Asymmetric Cascading Bandits for Multiple Clicks ​

Author: Andy Wang, Charlton Shih, William Chang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11873v1 Announce Type: new Abstract: In this work, we extend the Dependent Click Model (DCM) Bandits to a multiplayer information-asymmetric setting, where multiple agents interact with a shared ranked list and may observe multiple clicks per session, introducing new challenges for select...

📖 Read original article


69. Disentangling the Expressivity of RoPE ​

Author: Selim Jerad, Anej Svete, Jiaoda Li, Ryan Cotterell
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.FL

arXiv:2608.11909v1 Announce Type: new Abstract: Two accounts recur in explanations of the success of rotary position embeddings (RoPE). Expressivity studies associate periodic position information with modular predicates, whereas mechanistic and long-context studies emphasize positional anchors and ...

📖 Read original article


70. A Factor Graph Approach to Scalable Multi-Output Gaussian Process Regression ​

Author: Wouter W. L. Nuijten, Esther G. van Pelt, Albert Podusenko, .Ismail \c{S}en"oz, Wouter M. Kouw
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11917v1 Announce Type: new Abstract: Multi-output Gaussian process regression scales cubically in the number of observations times outputs, and dense kernel-matrix methods need bespoke handling whenever different outputs are observed at different inputs. We express multi-output Gaussian p...

📖 Read original article


71. Distillation of Foundation Models for Time-dependent PDEs ​

Author: Daniel Musekamp, Boshra Ariguib, Andrei Manolache, Mathias Niepert
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11937v1 Announce Type: new Abstract: Foundation models for time-dependent partial differential equations (PDEs) are trained on large and diverse collections of physical systems and can generalize effectively to new downstream tasks. After fine-tuning on only a few trajectories from a targ...

📖 Read original article


72. TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement ​

Author: Karim Aly, Alexei Sharpanskykh, Jacco Hoekstra
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11951v1 Announce Type: new Abstract: Extreme events in air transport, such as severe arrival delays and abnormal air times, cause cascading network disruptions with substantial operational, economic, and safety costs. Such events are rare in historical records, leaving insufficient traini...

📖 Read original article


73. LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation ​

Author: Zhixin Zhang, Xinke Jiang, Zhibang Yang, Weixuan Xu, Guohong Qiu, Xu Chu, Junfeng Zhao, Yasha Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11967v1 Announce Type: new Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical capability in such settings is reflection: assessing trajectory progress, identifying missing evidence a...

📖 Read original article


74. TESLA: Taylor Expansion of Sinusoidal Learnable Activations ​

Author: Daehwa Ko, Jaehyeon Kim, Seunghyun Ham, Jay Hoon Jung
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11970v1 Announce Type: new Abstract: The parity problem--deciding whether the number of ones in a binary vector is odd or even--remains challenging for standard neural networks due to linear inseparability and the need for global interactions. We propose TESLA, an activation defined as a ...

📖 Read original article


75. Remote Sensing and Machine Learning-Based Analysis of Land Use and Vegetation Change in Dhaka District, Bangladesh ​

Author: Muhammad Masud Tarek, Md. Alamgir Hossain, Md. Samiul Islam, Muntasir Hasan Kanchan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.12001v1 Announce Type: new Abstract: Rapid urbanization in Dhaka District, Bangladesh has triggered substantial alterations in land use and environmental conditions, necessitating systematic monitoring for informed urban planning and ecological sustainability. This study employs remote se...

📖 Read original article


76. Dual-Model Sentiment Analysis of Consumer Reviews in the Retail Coffee Sector Using Machine Learning and Deep Learning Approaches ​

Author: Muntasir Hasan Kanchan, Md. Alamgir Hossain, Md. Samiul Islam, Muhammad Masud Tarek
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.12007v1 Announce Type: new Abstract: Consumer reviews play an important role in shaping brand perception and business strategies, particularly in service-driven industries such as retail coffee. This study presents a comparative sentiment analysis framework for Starbucks customer reviews ...

📖 Read original article


77. Reducing Symmetry Increase in Equivariant Neural Networks ​

Author: Ning Lin, Jiacheng Cen, Anyi Li, Wenbing Huang, Hao Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12010v1 Announce Type: new Abstract: Equivariant Neural Networks (ENNs) have empowered numerous applications in scientific fields. Despite their remarkable capacity for representing geometric structures, ENNs suffer from degraded expressivity when processing symmetric inputs: the output r...

📖 Read original article


78. SoftWater: Class-Aware Rate Allocation for Softmax Quantization ​

Author: Joao V. Cavalcanti, Ashia C. Wilson
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12026v1 Announce Type: new Abstract: Post-training quantization pipelines routinely leave the softmax output layer in high precision. Yet in small LLMs with modern vocabularies, the head holds 15--30% of all parameters, so a nominal ``2-bit'' model with an fp16 head can store several tim...

📖 Read original article


79. Uncertainty-Aware Probabilistic Constrained Clustering from Entangled Pairwise Supervision ​

Author: Shaojie Zhang, Ke Chen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.12027v1 Announce Type: new Abstract: Pairwise constrained clustering typically relies on hard must-link/cannot-link labels, whereas realistic pairwise supervision may be real-valued and entangle intrinsic ambiguity, expert judgment, and stochastic corruption. Existing deep constrained clu...

📖 Read original article


80. Clustered Randomized Smoothing for Stochastic Prediction Functions ​

Author: Eduardo Figueiredo, Frederik Mathiesen, Julian Schumann, Jens Kober, Arkady Zgonnikov, Luca Laurenti
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2608.12037v1 Announce Type: new Abstract: Modern stochastic predictors can model rich, multi-modal outcome distributions. However, this expressive power comes with challenges in ensuring robust predictions $-$ a critical requirement in safety-critical domains. Randomized smoothing is a leading...

📖 Read original article


81. Towards Truly Unsupervised Evaluation of Feature Selection ​

Author: Hafiz Saud Arshad, Muhammad Rajabinasab, Arthur Zimek
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12057v1 Announce Type: new Abstract: Feature selection is one of the most important and fundamental tasks in data mining, tackled by a family of methods with an established set of evaluation techniques to measure the quality of a specific method. Most of the methods commonly used for the ...

📖 Read original article


82. Faithful, Sufficient and Understandable: Rethinking Graph Counterfactual Explanations via Discrete Diffusion Inversion ​

Author: David Bechtoldt, Sidney Bender
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.12083v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve strong predictive performance on graph-structured data across domains such as chemistry, biology, and network analysis, yet they provide no intrinsic explanation of their predictions. This limits their adoption in h...

📖 Read original article


83. NAE: Normalizing AutoEncoder ​

Author: Muhammad Abdur Rafae, Niels Landwehr
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12084v1 Announce Type: new Abstract: We consider the setting of Normalizing flows with approximate inverses, an established paradigm spanning both full-dimensional ($d=D$) and bottleneck ($d<D$) settings, and group these models under the term flow autoencoders. We present a theoretical in...

📖 Read original article


84. Task- and dataset-specific information in protein language models ​

Author: Roman Joeres, Ilya Senatorov, Anastasia Kolchina, Dietrich Klakow, Olga V. Kalinina
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM

arXiv:2608.12090v2 Announce Type: new Abstract: Protein language models (PLMs) have transferred the latest advances from natural language processing to computational biology. These models, trained on large corpora of protein sequence data, are widely used to translate amino acid sequences into laten...

📖 Read original article


85. Confidence Calibration of Deep Learning Systems ​

Author: Coby Penso
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.12100v1 Announce Type: new Abstract: In high-stakes applications, reliable confidence estimates are as important as the predictions themselves. Confidence calibration ensures that predicted probabilities reflect the likelihood of correctness, making it essential for safe deployment of dee...

📖 Read original article


86. Beyond Parameter Space: NTK-Guided Personalized Aggregation for Robust Federated Learning ​

Author: Mirko Konstantin, Stefan Zachow, Anirban Mukhopadhyay
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12108v2 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed clients while keeping data local. A central challenge is determining which client updates are beneficial for aggregation with respect to each client's target domain. Existi...

📖 Read original article


87. Attractor Image-Based Deep Learning of Arterial Pulse Waves for Age Classification ​

Author: Sara Vardanega, Patrick Segers, Philip Aston, Ernst Rietzschel, Jordi Alastruey, Manasi Nandi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12117v1 Announce Type: new Abstract: Arterial pulse waveform morphology evolves with age, reflecting structural and functional changes in the cardiovascular system. Thus, vascular age is a valuable surrogate marker of cardiovascular health, and premature vascular ageing can indicate incre...

📖 Read original article


88. Adversarial Resilience of Poisson-Process Submodular Maximization over Matroids: From Robust Offline Optimization to Full-Bandit Learning ​

Author: Vaneet Aggarwal
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CC, math.OC

arXiv:2608.12134v1 Announce Type: new Abstract: We study nonnegative submodular maximization subject to a general matroid when the offline algorithm is given an arbitrary controlled value oracle. Our main result is an adversarial resilience theorem for the Spiteful Greedy Swap Poisson Process (SGS-P...

📖 Read original article


89. HYDRA: Hyperbolic Dynamic Representation Architecture for Kolmogorov-Arnold Networks ​

Author: Zhao Su, Yuxin Xia, Haoran Li, Jun Shen, Qi Zhu, Qingguo Zhou, Binbin Yong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.12194v1 Announce Type: new Abstract: Kolmogorov-Arnold Networks (KANs) enhance nonlinear function approximation by replacing scalar weights with learnable univariate functions. However, assigning an independent function to every connection results in substantial parameter redundancy, limi...

📖 Read original article


90. ScreenShot: A Foundation Model for Few-Shot Combination Drug Screening ​

Author: Antoine de Mathelin, Christopher Tosh, Wesley Tansey
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12219v1 Announce Type: new Abstract: Treating patients with combinations of drugs reduces the risk of resistance to any individual drug. Finding effective combinations is difficult because the large search space makes combinatorial screens prohibitively expensive, time consuming, and ofte...

📖 Read original article


91. An Efficient Near-Optimal Algorithm for Adversarial $m$-Set Bandits ​

Author: Francesco Bacchiocchi, Tommaso Cesari, Roberto Colomboni
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12231v1 Announce Type: new Abstract: We study adversarial combinatorial bandits with $m$-set actions, where at each round the learner selects $m$ out of $d$ items and observes only the aggregate loss of the selected items. The resulting action set contains $K=\binom{d}{m}$ elements and ca...

📖 Read original article


92. Calibration Bets on the Past: Post-Training Quantization for Financial Time-Series Forecasting ​

Author: Junyi Ye, Ivy Gateri Wanjiku
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, q-fin.ST

arXiv:2608.12259v1 Announce Type: new Abstract: Financial forecasting models are typically developed in full precision, yet production deployment often requires low-precision inference to reduce memory and computational cost. Post-training quantization (PTQ) enables such deployment without retrainin...

📖 Read original article


93. Earth observation embeddings are effective sub-grid descriptors for probabilistic weather downscaling ​

Author: Pedro Sousa (Department of Computer Science, University of Cambridge), Will Tebbutt (Department of Engineering, University of Cambridge), Sadiq Jaffer (Department of Computer Science, University of Cambridge), Robin Young (Department of Computer Science, University of Cambridge), Anil Madhavapeddy (Department of Computer Science, University of Cambridge), Richard E. Turner (Department of Engineering, University of Cambridge)
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2608.12271v1 Announce Type: new Abstract: Global weather reanalyses and forecasts resolve the evolving atmospheric state on coarse grids, but site-specific applications require predictions at arbitrary locations where near-surface conditions also depend on unresolved terrain and land-surface p...

📖 Read original article


94. A Framework for Designing Reward Functions: From Objectives to Features to Human-Aligned Reward Functions ​

Author: Di Yang Shi, W. Bradley Knox
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12302v1 Announce Type: new Abstract: We present a formal process to enable non-experts to instantiate and iterate on human-aligned reward functions, i.e. reward functions that adhere to a given preference ordering over trajectories. Given a task described in natural language, our process ...

📖 Read original article


95. Redistribution-based Cost Inference Improves Sparse Safe Offline RL ​

Author: Ebenezer Gelo (University of the Witwatersrand), Geraud Nangue Tasse (University of the Witwatersrand), Steven James (University of the Witwatersrand), Benjamin Rosman (University of the Witwatersrand)
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.12306v1 Announce Type: new Abstract: Safe offline RL typically assumes access to dense per-step cost annotations, but in practice supervisors provide only trajectory-level stop-feedback: a binary signal at the first unsafe transition, with no per-step attribution. We frame this as a tempo...

📖 Read original article


96. AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses ​

Author: Cheng Qian, Wenting Zhao, Liangwei Yang, Heng Wang, Jielin Qiu, Heng Ji, Silvio Savarese, Huan Wang, Shelby Heinecke
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.12307v1 Announce Type: new Abstract: Recent work on distillation transfers the capabilities of large models to smaller ones often by updating the latter's parameters, through teacher forcing, on-policy distillation, and related training-time methods. In this paper, we ask whether such tra...

📖 Read original article


97. WavePhaseNet: A DFT-Based Method for Constructing Semantic Conceptual Hierarchy Structures (SCHS) ​

Author: Kiyotaka Kasubuchi, Kazuo Fukiya
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2602.14419v2 Announce Type: cross Abstract: This paper reformulates Transformer/Attention mechanisms in Large Language Models (LLMs) through measure theory and frequency analysis, theoretically demonstrating that hallucination is an inevitable structural limitation. The embedding space functio...

📖 Read original article


98. Reproducing and Stress-Testing Two Approaches to LLM Reasoning Reliability: Test-Time Probability Aggregation and Logic-Representation Editing ​

Author: Minhan Cho, Jimin Kweon
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.08514v1 Announce Type: cross Abstract: We independently reproduce two recent methods for making large language model (LLM) reasoning more reliable, and stress-test them across domains and models (RPC across four new task domains with Qwen3-8B, LCF across four 7-8B models). The first, RPC,...

📖 Read original article


99. What Iterated Self-Feeding Probes of Language Models Measure, and a test that separates the construction from the model ​

Author: Nicol'as Vera Z'u~niga
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.10986v1 Announce Type: cross Abstract: A growing class of methods probes a language model by feeding it its own output: self-consistency, iterated refinement, agentic loops. We ask what such a probe measures, in a construction chosen to make the question sharp: a ring of token cells resam...

📖 Read original article


100. Detecting a Route Flip Is Easier Than Knowing Whether to Fix It: Causal Route-Mediated Damage in Quantized Mixture-of-Experts ​

Author: Parvel Gu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.11212v1 Announce Type: cross Abstract: Top-k Mixture-of-Experts (MoE) routing is discontinuous, so a deployment-motivated numerical disturbance -- simulated 4-bit KV-cache quantization read by a protected BF16 gate -- pushes tokens across decision boundaries and flips which experts fire. ...

📖 Read original article


101. Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop ​

Author: Igor Itkin
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cond-mat.stat-mech, cs.CL, cs.LG, cs.MA, physics.soc-ph

arXiv:2608.11215v1 Announce Type: cross Abstract: Simulating societies of many large language model (LLM) agents is expensive, yet the questions asked of such simulations are usually macroscopic: phase behaviour, stylised facts, and scaling with the number of agents $N$, not the cognition of any sin...

📖 Read original article


102. MaSRead: Content-Addressed Reading of Replicated Latent Stores ​

Author: Carlos Baquero, Lu'is Brito, Jo~ao Resende
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2608.11218v1 Announce Type: cross Abstract: Independent agents that reason in latent space can share computed state as key-value cache fragments rather than text. Merged by a conflict-free replicated data type, these fragments form a store that converges under any delivery order or duplication...

📖 Read original article


103. From Monolithic to Modular: Segment-level Automatic Prompt Optimization ​

Author: Nikita Kulin, Viktor Zhuravlev, Artur Khairullin, Sergey Muravyov, Ilya Makarov, Daniil Sukhorukov, Ekaterina Averkova
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.11219v1 Announce Type: cross Abstract: Automatic Prompt Optimization (APO) often rewrites prompts monolithically, which can improve one behavior while degrading others. We present SAPO, a segment-level APO method that decomposes prompts into role, context, tasks, and output format, then a...

📖 Read original article


104. Transit Destination Inference from Tap-In-Only Bus Smart-Card Data: A Hierarchical Bayesian Approach ​

Author: Gefei Zhao, Jiahe Ling, Yuelong Su
Published: 8/13/2026, 4:00:00 AM
Categories: stat.AP, cs.LG, cs.SY, eess.SY, stat.ML

arXiv:2608.11223v1 Announce Type: cross Abstract: Entry-only automatic fare collection systems record boardings but not alightings, preventing direct construction of origin-destination (OD) matrices. This study develops a Hierarchical Bayesian Latent-Destination (HBLD) model that combines station-ho...

📖 Read original article


105. Forecasting Side Effects of Activation Steering ​

Author: Chong Yong Ong, Alson Wei Jie Sim, Peixin Zhang, Jun Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.11227v1 Announce Type: cross Abstract: Activation steering modifies a language model by adding a learned direction to its hidden activations, enabling targeted behavioral changes without retraining. While effective, steering often produces unintended side effects on other behaviors, makin...

📖 Read original article


106. Retrofitting Recurrent Depth into a Pretrained Language Model: Installation, Extrapolation, Transfer, and Retention at Two Parameter Budgets ​

Author: Mark Shapiro
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.11233v1 Announce Type: cross Abstract: A dense, pretrained language model can be retrofitted with recurrent depth and learn an iterative latent transition that persists after outcome-only annealing. Qwen2.5-0.5B-Instruct is split into a Prelude, a weight-tied Recurrent Block, and a Coda, ...

📖 Read original article


107. The Off-Support Barrier: Why Semantic Safety Constraints Are Not Learning-Problem Invariants, and What Follows for Prior Design, Containment, and Verification ​

Author: Yoshinori Watanabe
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.11243v1 Announce Type: cross Abstract: We argue that a single structural fact organizes a wide range of phenomena in contemporary AI safety: a semantic safety constraint (e.g., the agent does not escape its sandbox) is an off-support object. Formally, if q is the data distribution and (p...

📖 Read original article


108. Towards Sustainable Learning in Online Education: A Reinforcement Learning Approach ​

Author: Chaofan Zhai, Yicheng Song, Ravi Bapna, Junyao Ye
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG

arXiv:2608.11245v1 Announce Type: cross Abstract: Online education offers unprecedented scalability and accessibility to global learners from diverse backgrounds, but it often suffers from low engagement and poor long term learning effectiveness. To address these challenges, we introduce AI Tutor, a...

📖 Read original article


109. Towards the Harness of Embodied Agents ​

Author: Qi Wang, Tianyi Wang, Chengyang Li, Shikun Ban, Yurun Chen, Yizhong Ge, Jason Qin, Chengtai Li, Wentao Zhu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.RO

arXiv:2608.11246v1 Announce Type: cross Abstract: The success of coding agents has established the harness as a paradigm: what an agent achieves depends not on the model alone, but on the infrastructure around it. We ask whether the same paradigm extends to embodied agents in the physical world. We ...

📖 Read original article


110. Diffuse to Compress: Leveraging Diffusion LMs for Lossless Compression ​

Author: Angelo Nardone, Paolo Ferragina
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.IT, cs.LG, math.IT

arXiv:2608.11249v1 Announce Type: cross Abstract: We study the problem of lossless text compression, motivated by the rapid growth in the collection and storage of digital textual data - including plain text, source code, and structured formats such as XML - and by recent advances in neural language...

📖 Read original article


111. Variable Selection in the Context of AI Fairness ​

Author: Ivan Luciano Danesi, Chiara Frigerio, Fabio Maccaferri, Giorgio Alessandro Motta, Pietro Zecca
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG

arXiv:2608.11251v1 Announce Type: cross Abstract: Fairness in AI systems has become more important with recent regulatory demands, such as the EU AI Act. Traditional approaches often do not take into account philosophical ethics and social awareness. Variable selection processes, in particular, can ...

📖 Read original article


112. Symbolic Machine Learning for Vapor-Liquid Equilibrium Prediction in Cx-N2 Binary Mixtures ​

Author: Bongseok Kim, Suman Chakraborty, Gary Huang, Mehek Mathur, Guang Lin, Li Qiao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, physics.comp-ph

arXiv:2608.11255v1 Announce Type: cross Abstract: Accurate prediction of vapor--liquid equilibrium (VLE) for hydrocarbon-nitrogen mixtures remains challenging for cubic equations of state, particularly across broad ranges of composition and hydrocarbon chain length. While deep learning models can pr...

📖 Read original article


113. Temperature-Driven Sequential Modeling for the Prediction of Annual Power Conversion Efficiency Profiles of Organic Photovoltaic Materials: Douala Case Study ​

Author: Steve Cabrel Teguia Kouam, Rockefeller Rockefeller, Raoult Dabou Teukam, Jean-Pierre Tchapet Njafa, Patrick Sorrel Mvoto Kongo, Jean-Pierre Nguenang, Serge Guy Nana Engo
Published: 8/13/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG, physics.chem-ph

arXiv:2608.11261v1 Announce Type: cross Abstract: Organic photovoltaic (OPV) materials are promising candidates for distributed solar energy in tropical regions, yet existing virtual screening tools report static power conversion efficiency (PCE) values at standard testing conditions (STC) that fail...

📖 Read original article


114. CosMAP: Contrastive Manifold Approximation and Projection for Dimensionality Reduction of Omics and Genealogical Data ​

Author: Fenosoa Randrianjatovo, Maya Saleh, Simon Girard, Amadou Barry
Published: 8/13/2026, 4:00:00 AM
Categories: q-bio.GN, cs.LG, stat.CO, stat.ME

arXiv:2608.11269v1 Announce Type: cross Abstract: Omics datasets, particularly single-cell RNA sequencing data, are high-dimensional, sparse, noisy, and dominated by zero values, making faithful low-dimensional representation challenging. Existing dimensionality-reduction methods may distort local n...

📖 Read original article


115. Hardware-Aware Deployment of Joint SAR Compression and Despeckling on FPGA ​

Author: C'edric L'eonard, Francescopaolo Sica, Martin Schulz
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.LG

arXiv:2608.11271v1 Announce Type: cross Abstract: Next-generation Synthetic Aperture Radar (SAR) missions will generate data far faster than they can downlink, making onboard data reduction essential for near-real-time Earth observation. Learned Image Compression (LIC) offers better rate-distortion ...

📖 Read original article


116. Uncertainty-Aware and Explainable Ensemble Deep Learning Framework for Multi-Class Skin Lesion Classification ​

Author: Rofiqul Islam, Lilatul Ferdouse
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG

arXiv:2608.11280v1 Announce Type: cross Abstract: Skin cancer diagnosis from dermoscopic images remains challenging due to high intra-class variability, inter-class similarity, class imbalance, and the limited interpretability of deep learning models. This paper proposes an uncertainty-aware and exp...

📖 Read original article


117. Physics-Informed Implicit Neural Representations for Improved Myocardial Perfusion MRI Quantification ​

Author: Christos Tsepas, Chang Yan, Maximilian Fuetterer, Sebastian Kozerke, Cian M Scannell
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.LG

arXiv:2608.11282v1 Announce Type: cross Abstract: Quantifying myocardial perfusion from cardiac magnetic resonance (CMR) can be achieved by fitting tracer-kinetic models to the dynamic contrast-enhanced MR data. However, fitting the observed data with multi-compartment exchange models, which describ...

📖 Read original article


118. SegPAR: Class-Centric Decision-Based Sparse Attack for Semantic Segmentation ​

Author: Dongsu Song, DaeYun GO, Boseung Seo, Jay Hoon Jung
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CR, cs.LG

arXiv:2608.11285v1 Announce Type: cross Abstract: Despite the practical relevance of sparse decision-based black-box threats, they have received limited attention in semantic segmentation. To bridge this gap, we adapt the most representative decision-based black-box sparse attacks from the classific...

📖 Read original article


119. Benchmarking Cyberattack Detection in Electric Vehicle Charging Infrastructure with Benign User Updates ​

Author: Hannan Chen, Roshni Anna Jacob, Jie Zhang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.SY, eess.SY

arXiv:2608.11286v1 Announce Type: cross Abstract: Cyberattack detection in electric vehicle charging infrastructure is complicated by legitimate post-activation revisions to requested energy and departure time. Charging manipulation attacks can exploit the same interface and variables; therefore, de...

📖 Read original article


120. CLEAR: Class-wise Expert Aggregation with Structured Sampling for Long-Tailed Classification ​

Author: Gawon Lim
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.11287v1 Announce Type: cross Abstract: Long-tailed classification poses a reliability challenge because models trained on imbalanced data are unevenly reliable across frequent and underrepresented classes. While existing methods address imbalance through re-balancing, adjustment, represen...

📖 Read original article


121. Uncertainty-Aware Compositional Localization and Placement Assessment of Catheters and Tubes in Chest X-Rays ​

Author: Harshil Lodhiya
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2608.11288v1 Announce Type: cross Abstract: Assessing catheter and tube placement on chest X-rays is safety-critical yet tedious and error-prone. Current deep learning methods either classify placement globally -- losing track of which device is where -- or segment all devices into a single ma...

📖 Read original article


122. Dueling Deep Q-Learning for Intrusion Detection ​

Author: Logan Luna (Georgia Institute of Technology), Matthew P. Berkowitz (Embry-Riddle Aeronautical University), Laxima Niure Kandel (Embry-Riddle Aeronautical University), Sirio Jansen-S'anchez (Embry-Riddle Aeronautical University)
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.11291v1 Announce Type: cross Abstract: Intrusion detection systems (IDS) and automated systems for detecting and reporting cyber threats, are commonly handled via supervised machine learning methods. Though effective, these models struggle to effectively adapt to new attack types. This st...

📖 Read original article


123. Clinical Feasibility of Low-Magnification Fluorescence Imaging for Breast Cancer Margin Detection Using Texture Analysis and Deep Learning ​

Author: Pouya Afshin, Tianling Niu, Tongtong Lu, David Helminiak, Julie Jorns, Mollie Patton, Tina Yen, Donghye Ye, Bing Yu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.11317v1 Announce Type: cross Abstract: High-resolution images of unprocessed surgical breast tissue can be obtained using microscopy with ultraviolet surface excitation (MUSE). This technique is considered a promising method for checking surgical margins during breast cancer surgery. In t...

📖 Read original article


124. Spectral graph clustering with inhomogeneous latent geometry ​

Author: Konstantin Avrachenkov, Lucas S. Sibemberg, Alexander Van Werde
Published: 8/13/2026, 4:00:00 AM
Categories: cs.SI, cs.LG, math.PR, stat.ML

arXiv:2608.11321v1 Announce Type: cross Abstract: We study spectral clustering in the presence of a confounding latent geometry. The leading eigenvectors may then be dominated by the latent geometry rather than by the communities. Nevertheless, we show in a block latent-space model that communities ...

📖 Read original article


125. Deployment Decision Reliability: A Generalizability-Theory Framework for Sizing Long-Horizon Agent Evaluations ​

Author: Vasundra Srinivasan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.11323v1 Announce Type: cross Abstract: Enterprise practitioners read agent leaderboards as if they ranked agent capability. We show, across three open agent-trace benchmarks (TheAgentCompany, $\tau^2$-bench, and AppWorld), that the agent main effect accounts for less than 3% of total vari...

📖 Read original article


126. Better, Faster, Stronger: Programmatic Skill Learning Best Reduces Agent Cost ​

Author: Zixi Huang, Xiheng Wang, Andrew Wang, William Jurayj, Bernal Jim'enez Guti'errez, Daniel Khashabi, Nicholas Andrews
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.11338v1 Announce Type: cross Abstract: Recently, the practice of augmenting LLM agent capability with skills has gained prevalence. We explore the cost effective adaptation of agents to novel domains by means of learning skills. Existing works focus on performance gain over cost effective...

📖 Read original article


127. Can Frontier LLMs Match Natively Multimodal Embeddings? A Comparison on Hard-Negative Text-to-Image Retrieval ​

Author: Archan Dutta, Vyanktesh Kanungo
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.IR, cs.LG

arXiv:2608.11343v1 Announce Type: cross Abstract: Multimodal retrieval and classification across different types of media, spanning text, images,video and audio, has traditionally relied on dual-encoder models that align visual and textual representations through contrastive learning. The March 2026...

📖 Read original article


128. ODE-Based Transformer Decoders for Iterative Sign Language Translation ​

Author: Tu\u{g}\c{c}e K{\i}z{\i}ltepe, Hacer Yalim Keles
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.11352v1 Announce Type: cross Abstract: Sign language translation has achieved strong results with Transformer architectures, yet recent improvements largely rely on scaling model capacity at the cost of increased computation. We propose a parameter-efficient alternative that improves expr...

📖 Read original article


129. Adaptation of Generalist Robot Policies with Minimal Data ​

Author: Shreyas Kowshik, Sreyas Venkataraman, Leo Wang, Niharika Pant, Max Simchowitz, Aviral Kumar
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.11363v1 Announce Type: cross Abstract: A central goal in robot learning is to move beyond task-specific human data collection toward robots that improve through autonomous interaction. Yet fully autonomous learning remains difficult with current policies: sparse rewards and weak zero-shot...

📖 Read original article


130. From Numbers to Judgment: Specialist LLM Agents and Reinforcement Learning for European Listed Real Estate ​

Author: Pardis Taghavi, Santosh Bhavani
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.11381v1 Announce Type: cross Abstract: We study whether the localized numerical operations and integrative judgments of financial analysis benefit from the same form of LLM specialization. Larix maps a 16-lens European listed-real-estate analysis framework to eight lens-aligned specialist...

📖 Read original article


131. Generative Learning for Quantum Measurement Design ​

Author: Jun Dai, Olivier Nahman-L'{e}vesque, Guillaume Rabusseau, Hong-Ye Hu, Cunlu Zhou
Published: 8/13/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.11396v1 Announce Type: cross Abstract: Extracting quantum information from a quantum state is a fundamental task of quantum computation, often requiring the estimation of many non-commuting observables under a finite measurement budget. For both near-term and early fault-tolerant settings...

📖 Read original article


132. When Self-Consistency Backfires: Majority Vote Hurts the Majority of Hard Science Problems for Small LLMs ​

Author: Utkarsh Bahuguna
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.11403v1 Announce Type: cross Abstract: Self-consistency (SC) via majority vote is a widely used way to spend inference-time compute: sample N chains of thought, return the plurality answer. On the full GPQA Diamond benchmark (198 graduate-level science questions), majority voting reduces ...

📖 Read original article


133. Large-scale AI-Ready Data for Anti-Cancer Drug Response Modeling ​

Author: Vincent Lavelle, Yitan Zhu, Kaitlyn Marlor, Thomas Brettin, Rick Stevens
Published: 8/13/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG

arXiv:2608.11444v1 Announce Type: cross Abstract: Drug response prediction (DRP) models are an active area of research in pharmacogenomics, with growing potential to accelerate the identification of effective anticancer drugs. However, their predictive performance is often constrained by limited dat...

📖 Read original article


134. Gaussian Meta-Space Augmentation for Stacking Ensembles in Multimodal IPMN Risk Stratification ​

Author: Max A. Nelson, Eminenur Sen Tasci, Zhixiang Wang, Zongwei Zhou, Halil Ertugrul Aktas, Andrea M. Bejar, Elif Keles, Ziliang Hong, S{\i}tk{\i} Safa Taflan, Muhammed Enes Tasci, Frank H. Miller, Michael B. Wallace, Rajesh N. Keswani, Gorkem Durak, Ulas Bagci
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.11472v1 Announce Type: cross Abstract: Pancreatic cancer is among the most lethal malignancies; risk stratification of intraductal papillary mucinous neoplasms (IPMNs) offers a crucial opportunity for early intervention but typically requires invasive tissue biopsy. Dominant vision-based ...

📖 Read original article


135. Probing and steering biology across Boltz-1s trunk-diffusion boundary ​

Author: Piotr Jedryszek, Tongmeng Xie, Adam Winnifrith, Alexander Hasson, Weronika 'Slesak, George Wicks, Toby Winnifrith, Oliver M. Crook
Published: 8/13/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG

arXiv:2608.11475v1 Announce Type: cross Abstract: AlphaFold3-class structure predictors pair a representational trunk, which processes sequence and context, with a diffusion module, which generates atomic coordinates. How biological information changes as it crosses this architectural boundary remai...

📖 Read original article


136. Forward Trajectory Steering for Hamilton-Jacobi Reachability Analysis ​

Author: Sungje Park, Stephen Tu
Published: 8/13/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2608.11480v1 Announce Type: cross Abstract: Hamilton-Jacobi (HJ) reachability provides a mathematically rigorous framework for safe control of dynamical systems, but its practical application is bottlenecked by the computational complexity of solving Hamilton-Jacobi-Isaacs variational inequali...

📖 Read original article


137. A Modular Agentic Framework for Synthetically Constrained Multi-Objective Hit-to-Lead Optimization ​

Author: Kelvin P. Idanwekhai, Enes Kelestemur, Benjamin Strickland, Matthew Hart, Steini Davidsson, Angelos Angelopoulos, Ron Alterovitz, Marcello DeLuca, Alexander Tropsha
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-bio.QM

arXiv:2608.11483v1 Announce Type: cross Abstract: Hit-to-lead optimization requires iterative design of hit analogs across competing potency, selectivity, physicochemical, pharmacokinetic, safety, and synthetic constraints. We present SABLE (Synthetically-accessible Agentic Bayesian Ligand Explorati...

📖 Read original article


138. Cross-Corpus Evaluation of Generalizable Vulnerability Detection in IoT Firmware ​

Author: Sadib Hassan Rumman, Md. Shariful Islam, Md. Rayhanur Rahman
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.11492v1 Announce Type: cross Abstract: IoT firmware vulnerability detection remains challenging due to heterogeneous firmware ecosystems, resource-constrained platforms, and limitations in existing benchmarks. Many datasets are synthetic or general-purpose and lack human-verified, contami...

📖 Read original article


139. From Prompting to Behavioral Alignment: Personalized LLM Judges for Recommendation Evaluation ​

Author: Alireza S. Ziabari, Kat Ellis, Colleen Chan, Ding Tong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.11493v1 Announce Type: cross Abstract: Traditional offline recommendation evaluation relies heavily on complex, manually maintained feature pipelines that are difficult to scale. While Large Language Models (LLMs) offer a promising alternative by predicting user engagement directly from r...

📖 Read original article


140. Language-Structured Relational Q-Learning for Threat-Aware Control in Safety-Critical Driving ​

Author: Aditya Humnabadkar, Huaizhong Zhang, Ardhendu Behera
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.ET, cs.LG

arXiv:2608.11498v1 Announce Type: cross Abstract: Natural-language-based scenario generation offers an intuitive means of describing rare and complex driving interactions, yet it is still uncertain whether training with language-structured data leads to truly adaptive control policies. We propose La...

📖 Read original article


141. Generative Semantic Segmentation via an Observable Semantic-Image Interface and Hierarchical Generator Evidence Alignment ​

Author: Weize Cai, Yongqi Dong, Zhida Shao, Zixin Fu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV

arXiv:2608.11537v1 Announce Type: cross Abstract: Generative semantic segmentation exposes structured predictions as images, but direct color decoding is susceptible to color drift and boundary mixing, whereas latent-feature decoders that predict a separate output distribution may relegate the rende...

📖 Read original article


142. Fine-Tuning Generative Models for Extreme Events via CVaR-Penalized Wasserstein Gradient Flows ​

Author: Thejani Gamage, Hyemin Gu, Zhizhen Zhang, Ziyu Chen, Markos Katsoulakis, Luc Rey-Bellet
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.11544v1 Announce Type: cross Abstract: We propose CVaR-penalized Generative Particle Algorithm (CVaR-GPA), a robust, tail-agnostic algorithm for fine-tuning generative models to learn heavy-tailed distributions and capture extreme events, requiring no prior knowledge or estimation of the ...

📖 Read original article


143. Beyond Single-Turn Confidence: Trajectory-Adapted Uncertainty Quantification for LLM Agents ​

Author: Dylan Bouchard, Mohit Singh Chauhan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.11552v1 Announce Type: cross Abstract: Uncertainty quantification (UQ) methods for language models are typically evaluated on single-turn outputs, where uncertainty is attached to one generated answer. For LLM agents, however, the unit of observation is an interactive trajectory, where th...

📖 Read original article


144. Unifying Physical Backpropagation ​

Author: Cyrill B"osch, Yigithan Gediz, Hakan T"ureci
Published: 8/13/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cs.ET, cs.LG, physics.optics

arXiv:2608.11585v1 Announce Type: cross Abstract: Physical computing systems exploit device dynamics for computation, but their gradient-based optimization is challenging: backpropagation through a digital twin suffers from model-reality gap. On-device gradient computation could resolve this issue, ...

📖 Read original article


145. Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning ​

Author: Xulin Fan, Jialu Li, Mohammad Nur Hossain Khan, Kexin Hu, Bashima Islam, Mark Hasegawa-Johnson, Nancy L. McElwain
Published: 8/13/2026, 4:00:00 AM
Categories: eess.AS, cs.CL, cs.LG

arXiv:2608.11587v1 Announce Type: cross Abstract: Recent advances in model design and self-supervised audio representations have improved speech and audio understanding, yet infant-centered naturalistic recordings remain challenging due to limited labeled data, low signal-to-noise ratio, and cross-f...

📖 Read original article


146. CookVoice: Unified Framework for Style Controllable Multi-Modal Human Voice Generation ​

Author: Haowei Lou, Hye-Young Paik, Dai Jia, Kai Li, Lina Yao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.SD, cs.LG

arXiv:2608.11590v2 Announce Type: cross Abstract: Human voice generation has made rapid progress in speech generation, singing voice generation, voice cloning, and voice editing. However, most existing systems are designed for specific tasks and often rely on task-dependent architectures, control si...

📖 Read original article


147. IoT-Enabled Autonomous Maritime Navigation in Smart Ports: A Curriculum-Guided Shared Policy Learning Framework ​

Author: Yuqing Lin, Rangya Zhang, Kum Fai Yuen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.11597v1 Announce Type: cross Abstract: As smart port infrastructures increasingly rely on autonomous maritime devices enabled by the Internet of Things (IoT), ensuring reliable onboard navigation intelligence has become a critical challenge for safe and scalable operations in congested wa...

📖 Read original article


148. MBA: Multimodal Benchmark and Agents for Real-World Business Ideation ​

Author: Hojun Choi, Jaeyo Shin, Suin Lee, Hyunjung Shim
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG

arXiv:2608.11616v2 Announce Type: cross Abstract: Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation. Yet existing approaches remain confined to a text-only paradigm, despite the inherently multimodal nature of real-world contexts. We thus int...

📖 Read original article


149. CAM-Guided Saliency Cutout and Image-Based Malware Classification ​

Author: Yasaman Ebrahimi, Martin Jurecek, Mark Stamp
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.11634v1 Announce Type: cross Abstract: Dropout regularization is commonly used to reduce overfitting by removing parts of a neural network during training. For Convolutional Neural Networks (CNN), cutouts serve a somewhat analogous purpose. Cutouts can be implemented as data augmentation:...

📖 Read original article


150. Robustness of AI-Art Detectors under Generator Shift ​

Author: Shivank Singh Thakur, Meien Li, Mark Stamp
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.11643v1 Announce Type: cross Abstract: Text-to-image generative models have advanced rapidly, with modern Diffusion Transformer architectures producing images that are increasingly difficult to distinguish from human-created artwork. This development has raised significant concerns regard...

📖 Read original article


151. A Quantum/Classical Example Oracle Separation for Making Things Up ​

Author: Kenny Chen
Published: 8/13/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, stat.ML

arXiv:2608.11648v1 Announce Type: cross Abstract: We study the power of quantum examples, as compared to classical examples, in the PAC learning framework. Here, we have two learning algorithms, both with access to quantum computation, but one gets quantum examples, whereas the other gets classical ...

📖 Read original article


152. Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing ​

Author: Tianci Liu, Zihan Dong, Tianchun Li, Yi-Chung Chen, Qiming Cao, Xingchen Wang, Shiyang Wang, Zichen Miao, Linjun Zhang, Haoyu Wang, Jing Gao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.11660v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable performance across natural language tasks, yet they are trained on static corpora and their knowledge quickly becomes outdated in a fast-changing world. This motivates knowledge editing (KE), which upda...

📖 Read original article


153. APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference ​

Author: Alish Kanani, Layan Badawi, Umit Y. Ogras
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AR, cs.AI, cs.LG

arXiv:2608.11688v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models are attractive for edge deployment because they provide high model capacity while activating only a small subset of parameters per token, improving compute efficiency. However, MoE inference at the edge is fundamentall...

📖 Read original article


154. Locating and Controlling Implicit Personalization in Large Language Models ​

Author: Yueru Yan, Siqi Wu, Thai Le
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.11735v1 Announce Type: cross Abstract: Large language models (LLMs) often shift their outputs in response to implicit demographic cues even when users never state a demographic identity. Previous work has documented this behavior, but the connection between these behavioral changes and th...

📖 Read original article


155. Automated binary classification of hazelnut X-ray images: A deep-learning benchmark for quality assessment ​

Author: Giancarlo Sportelli, Nicola Belcari, Roberta Pace, Umberto Bernardo, Sharmin Sultana, Alessandra Toncelli, Matteo Giaccone
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, physics.app-ph

arXiv:2608.11759v1 Announce Type: cross Abstract: Non-destructive X-ray imaging can reveal internal hazelnut defects that are difficult to detect by external inspection alone; however, automated interpretation remains challenging because of subtle radiographic differences among classes, marked class...

📖 Read original article


156. Tight Nonasymptotic Local Convergence of Sinkhorn-Knopp ​

Author: Wenzhi Gao, Zhaonan Qu, Yinyu Ye, Madeleine Udell
Published: 8/13/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ML

arXiv:2608.11760v2 Announce Type: cross Abstract: We revisit the Sinkhorn-Knopp (SK) algorithm for the matrix scaling problem. Despite extensive literature on the global convergence of SK and its variants, its local linear convergence behavior remains less understood. We address this gap by providin...

📖 Read original article


157. A comparison of CNN architectures for Alzheimer's disease detection in single-view MRI scans ​

Author: Hiram Zuniga, Ulises Orozco-Rosas, Kenia Picos
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2608.11762v1 Announce Type: cross Abstract: Alzheimer's disease is a leading cause of death with no cure. Therefore, early detection is critical to slow progression and preserve quality of life. Diagnosis relies on medical history, cognitive tests, physical exams, and MRI brain scans, making d...

📖 Read original article


158. Achieving Near-Zero-Overhead Multi-Model Hierarchical Classification in Real-Time Detection Pipelines ​

Author: Vaishnav Raju
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.DC, cs.LG

arXiv:2608.11770v1 Announce Type: cross Abstract: Edge-deployed vision systems in target recognition, surveillance, autonomous vehicles, and drone domains require hierarchical inference pipelines where a detection model identifies objects of interest and downstream classifiers provide fine-grained a...

📖 Read original article


159. GRPO for Financial Advice Generation: Outperforming Commercial LLMs under CATE Evaluation ​

Author: Ofir Ben Shoham, Shrutendra Harsola, Vignesh Subrahmaniam, Shravan Mohan, Yakov Gazman, Oded Vainas
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.11787v1 Announce Type: cross Abstract: Generating actionable financial advice from business records demands that models integrate numerical reasoning, domain knowledge, and sound judgment, while avoiding recommendations that could harm the business. Direct supervision is difficult: histor...

📖 Read original article


160. Can Vision Models Read the Radar Display? On the Feasibility of Radar Imagery for Air Traffic Complexity Estimation ​

Author: Hyewook Kim, Byul Kang, Seokbin Yoon, Keumjin Lee
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.11810v1 Announce Type: cross Abstract: Air traffic controllers perceive traffic complexity through the radar display, suggesting that a computer vision model operating on the same imagery may provide a natural architecture for modeling controller-perceived complexity; however, whether rad...

📖 Read original article


161. Located but Not Releasable: Silent Gate Inversion and Bounded Linear Release ​

Author: Xining Xun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.11822v1 Announce Type: cross Abstract: A growing body of work reports that language models represent task-relevant latent structure that they fail to use. Whether such structure, once located, can be converted into behavior is a separate question that is rarely tested end to end. We submi...

📖 Read original article


162. User-Assisted Collaborative Distributed Inference for Efficient QoS-Aware Autoscaling ​

Author: Alfreds Lapkovskis, Ali Beikmohammadi, Sindri Magn'usson, Praveen Kumar Donta
Published: 8/13/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG, cs.NI, cs.PF

arXiv:2608.11840v1 Announce Type: cross Abstract: Growing demand for artificial intelligence (AI) inference services requires scalable infrastructure, yet centralized serving costs rise with demand. We propose a collaborative distributed inference system combining dedicated infrastructure with resou...

📖 Read original article


163. Policy-as-logic for robust reasoning over rules ​

Author: Rahul Nair, Bastian Lipka, Elizabeth Daly
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SC

arXiv:2608.11905v1 Announce Type: cross Abstract: In many practical applications of generative AI systems, from tax rules to airline baggage allowance, responses to natural language queries must respect written policies or rules. We present a hybrid symbolic approach that expresses policies in forma...

📖 Read original article


164. LODESTAR: Trustworthy Entropy Is Navigated, Not Merely Measured -- Reinforced Polarizer Keeps a Frozen LLM from Being Confidently Misled by the Wrong Evidence ​

Author: Po-Jen Ko, Che-Cheng Wu, Hung-Chun Hsu, Li-Yang Chang, Chuan-Ju Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.IR, cs.LG

arXiv:2608.11922v1 Announce Type: cross Abstract: Predictive-distribution entropy makes a strong selection rule in retrieval-augmented question answering: across five QA benchmarks, keeping the candidate answer that a frozen respondent LLM produces with the lowest answer-token entropy lifts mean ans...

📖 Read original article


165. Latent variable models for simultaneous EOV identification and removal in population-based SHM ​

Author: M. D. Champneys, M. R. Jones, A. J. Hughes, T. J. Rogers, E. J. Cross, K. Worden
Published: 8/13/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.11995v1 Announce Type: cross Abstract: The robust treatment of environmental and operational variability (EOV) is an open challenge in population-based structural health monitoring (PBSHM). The difficulty is compounded in the case that the EOV signals are unmeasured. A common approach in ...

📖 Read original article


166. A Remote Approach to Cashew Orchard Detection: Leveraging Active Learning with Satellite Imagery in Guinea-Bissau ​

Author: Miguel, Sofia, Maria, Patr'icia, Luke, Jo~ao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.11996v1 Announce Type: cross Abstract: Cashew production is a widespread economic activity in Guinea-Bissau, as well as other countries in West Africa. However, unregulated cashew production can be directly associated with increasing regionwide deforestation rates, biodiversity losses, an...

📖 Read original article


167. Beyond Local Power: Functional Connectivity Analysis for Subject-Independent Learning Style Recognition ​

Author: Wiga Maulana Baihaqi, Indriana Hidayah, Sri Kusrohmaniah, Noor Akhmad Setiawan
Published: 8/13/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG, eess.SP

arXiv:2608.12000v1 Announce Type: cross Abstract: Identifying individual learning styles optimizes pedagogical efficacy. While traditional questionnaires are structured, behavioral tracking methods require prolonged interaction log accumulation. To overcome these temporal constraints, this paper pro...

📖 Read original article


168. Adaptive Bregman Proximal Stochastic Gradient with a Stabilized Barzilai--Borwein Step Size ​

Author: Chenhan Jin, Shengze Xu, Binghui Xie, Kaiwen Zhou, Fan Jia, James Cheng, Tieyong Zeng
Published: 8/13/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2608.12009v1 Announce Type: cross Abstract: Bregman proximal stochastic gradient (BPSG) methods bring variance-reduced composite optimization to objectives whose geometry is poorly captured by Euclidean smoothness. Their performance, however, remains sensitive to the step size: raw stochastic ...

📖 Read original article


169. Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence ​

Author: Mengru Wang, Junfeng Fang, Shuofei Qiao, Zhenqian Xu, Haoming Xu, Haoxiong Wang, Shumin Deng, Linyi Yang, Zhixiang Cui, Xin Xu, Yunzhi Yao, Buqiang Xu, Fei Shen, Haozhe Luo, Yunxiang Wei, Ningyu Zhang, Julian McAuley, Tat Seng Chua, Huajun Chen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.HC, cs.LG, cs.MA

arXiv:2608.12036v1 Announce Type: cross Abstract: AI models have achieved remarkable success across diverse domains, yet the mechanisms underlying their capabilities and the risks they may pose remain poorly understood. As AI development becomes faster and increasingly automated, mechanistic explora...

📖 Read original article


170. Direct Acceleration of Stochastic Root-Finding Without Variance Reduction and Regularization ​

Author: TaeHo Yoon, Nicolas Loizou
Published: 8/13/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2608.12043v1 Announce Type: cross Abstract: Acceleration for deterministic root-finding problems has been extensively studied in recent years; specifically, the anchor-based, or Halpern-type methods achieve optimal convergence rates with respect to the operator norm. However, acceleration via ...

📖 Read original article


171. Draw This First ​

Author: Dazhi Zhong, Rowan Bradbury, Grant Davis
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.12064v1 Announce Type: cross Abstract: We invert the typical formulation of sketch generation: instead of drawing strokes in order, we predict a 2D field that defines the order in which strokes are drawn. We use a pretrained latent flow-matching transformer to supply the image prior to pr...

📖 Read original article


172. Better Slots, Better Worlds: Representation Quality & Robustness in Object-Centric World Models ​

Author: Shukrullo Nazirjonov, Sai Prasanna, Anna Manasyan, Georg Martius
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.12078v1 Announce Type: cross Abstract: Learning world models from offline trajectories enables agents to accomplish different tasks through planning. Object-centric (OC) representations, which decompose a scene into a set of slots that bind to its objects, have been proposed as an inducti...

📖 Read original article


173. Look What the Probes Dragged In! Real-World Chest X-ray Shortcuts in MedCLIP ​

Author: Nikolette Pedersen, Regitze Sydendal, Veronika Cheplygina, Th'eo Sourget
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.12086v1 Announce Type: cross Abstract: Vision-language models, such as contrastive language-image pre-training (CLIP)-based approaches, have reached state-of-the-art (SOTA) results in medical artificial intelligence. However, recent work reveals that CLIP-based models remain vulnerable to...

📖 Read original article


174. The Advective Fisher-Rao Geometry of Deterministic Measure Transport ​

Author: Benjamin Gess, Johannes M"uller
Published: 8/13/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DG, math.PR

arXiv:2608.12111v1 Announce Type: cross Abstract: A novel advective Fisher-Rao metric is introduced for optimization tasks on paths of probability measures governed by the continuity equation. This metric is shown to lead to optimal descent directions. It is then shown that this metric arises natura...

📖 Read original article


175. A corpus-specific clinical RAG system matches or outperforms newer frontier LLMs on HealthBench ​

Author: Praveen Reddy, Charuta Mandke, Suvrankar Datta, Sarah Khan, Siddharth Reddy Anthireddy, Shitij Arora, Vishal Singh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC, cs.IR, cs.LG

arXiv:2608.12138v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) have recently been reported to match or exceed specialized clinical AI tools on medical benchmarks, but such comparisons draw on a narrow set of systems and on benchmarks developed largely in high-income s...

📖 Read original article


176. FQTree: Fine-grained Quantization and Hardware Generation of Boosted Decision Trees ​

Author: Zhiqiang Que, Chang Sun, Haiyang Wang, Dinesh Pamunuwa, Roshan Weerasekera, Qijia Tang, Bakhtiar Zadeh, Wayne Luk, Maria Spiropulu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AR, cs.LG

arXiv:2608.12140v1 Announce Type: cross Abstract: Boosted decision trees (BDTs) are widely used in latency-critical applications, but efficient hardware deployment remains challenging. Existing designs often rely on uniform or manually tuned fixed-point formats, which can introduce unnecessary hardw...

📖 Read original article


177. ADEPT: A Unified Framework for Deep Learning Test Adequacy ​

Author: Yidi Kao, Shawn Burnham, Tommi Rose Fahy, Ali Ghanbari
Published: 8/13/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2608.12144v1 Announce Type: cross Abstract: Over the past decade, many test adequacy metrics have been proposed for deep learning that characterize test dataset adequacy from different perspectives, e.g., neuron activation behavior, latent feature coverage, decision-boundary exploration, etc. ...

📖 Read original article


178. Autonomous Telerehabilitation via Skeletal Motion Prediction and Joint-Level Performance Assessment ​

Author: Lara Pereira, Jo~ao Ruivo Paulo, Pedro Santos, Paulo Peixoto
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.12145v1 Announce Type: cross Abstract: Autonomous rehabilitation systems must not only recognize human motion but also provide structured feedback to support users without continuous therapist supervision. This paper presents a telerehabilitation pipeline that integrates skeleton-based ex...

📖 Read original article


179. How to Spend Your Oracle Budget: Practical Guidance for Protein Structure Prediction Models ​

Author: Aleksandra Kalisz, Jack Simons, Krisztina Sinkovics, Noam Ghenassia, Shikha Surana, Henry Moss, Paul Duckworth
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.12192v1 Announce Type: cross Abstract: Foundation models for protein structure prediction remain unreliable on certain targets. External oracles can flag and correct these failures, but biological oracles are expensive, making oracle budget a critical constraint. Existing guidance methods...

📖 Read original article


180. Learning-Based Behavior Planning for Automated Driving: Real-World Integration and Deployment ​

Author: Jean-Pierre Busch, Guido Linden, Jan Bergmann, Lutz Eckstein
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2608.12198v1 Announce Type: cross Abstract: Recent research in machine and deep learning has shown the potential of learningbased motion planning approaches to improve the driving behavior of automated vehicles, especially in complex environments. However, their complex nature and lack of tran...

📖 Read original article


181. Regime-Gated Residual Mixture-of-Experts for Cross-Sectional Volatility Forecasting ​

Author: Junyi Ye, Gargi Vijay Borde
Published: 8/13/2026, 4:00:00 AM
Categories: q-fin.ST, cs.LG

arXiv:2608.12251v1 Announce Type: cross Abstract: Financial volatility is regime dependent, yet incorporating regime information into neural networks can also destabilize training. This paper asks where such information should enter a neural cross-sectional volatility forecasting model. We study fiv...

📖 Read original article


182. One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL ​

Author: Simon Yu, Nicholas Tomlin, Marwa Abdulhai, Ximing Lu, Derek Chong, Abe Hou, Dilara Soylu, Sergey Levine, Christopher D. Manning, Weiyan Shi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that this approach systematically fails to generalize, and trace the failure to simulator collapse: becau...

📖 Read original article


183. Weak Correlations as the Underlying Principle for Linearization of Gradient-Based Learning Systems ​

Author: Ori Shem-Ur, Khen Cohen, Aviv Orly, Yaron Oz
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, hep-th, math.PR, stat.ML

arXiv:2401.04013v3 Announce Type: replace Abstract: Deep learning models, such as wide neural networks, can be conceptualized as nonlinear dynamical physical systems characterized by a multitude of interacting degrees of freedom. Such systems in the infinite limit, tend to exhibit simplified dynamic...

📖 Read original article


184. Deep Activity Model: A Generative Approach for Human Mobility Pattern Synthesis ​

Author: Xishun Liao, Qinhua Jiang, Brian Yueshuai He, Yifan Liu, Chenchen Kuai, Jiaqi Ma
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2405.17468v3 Announce Type: replace Abstract: Human mobility plays a crucial role in transportation, urban planning, and public health, but current approaches face important limitations. Existing deep learning models tend to overlook the semantic interdependencies among activities and househol...

📖 Read original article


185. A New First-Order Meta-Learning Algorithm with Convergence Guarantees ​

Author: El Mahdi Chayti, Martin Jaggi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2409.03682v2 Announce Type: replace Abstract: Learning new tasks by leveraging prior experience is a fundamental trait of intelligent systems. While Model-Agnostic Meta-Learning (MAML) is a leading approach, it suffers from significant computational and memory overhead due to the requirement o...

📖 Read original article


186. Program Semantic Inequivalence Game with Large Language Models ​

Author: Antonio Valerio Miceli-Barone, Vaishak Belle, Ali Payani
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.PL

arXiv:2505.03818v3 Announce Type: replace Abstract: Large Language Models (LLMs) can achieve strong performance on everyday coding tasks, but they can fail on complex tasks that require non-trivial reasoning about program semantics. Finding training examples to teach LLMs to solve these tasks can be...

📖 Read original article


187. Computational Algebra with Attention: Transformer Oracles for Border Basis Algorithms ​

Author: Hiroshi Kera, Nico Pelleriti, Yuki Ishihara, Max Zimmer, Sebastian Pokutta
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.SC

arXiv:2505.23696v2 Announce Type: replace Abstract: Solving systems of polynomial equations, particularly those with finitely many solutions, is a crucial challenge across many scientific fields. Traditional methods like Gr"obner and Border bases are fundamental but suffer from high computational c...

📖 Read original article


188. Improving Performance of Spike-based Deep Q-Learning using Ternary Neurons ​

Author: Aref Ghoreishee, Abhishek Mishra, John Walsh, Anup Das, Nagarajan Kandasamy
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, cs.SY, eess.SY

arXiv:2506.03392v3 Announce Type: replace Abstract: We propose a new ternary spiking neuron model to improve the representation capacity of binary spiking neurons in deep Q-learning. Although a ternary neuron model has recently been introduced to overcome the limited representation capacity offered ...

📖 Read original article


189. Learning Multi-Timescale Interventions under Safety and Resource Constraints ​

Author: David Mguni, Wanrong Yang, Jing Dong, Jing Peng, Ziquan Liu, Muhammad Salman Haleem, Baoxiang Wang, Dominik Wojtczak
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2508.03875v3 Announce Type: replace Abstract: Many sequential decision problems offer qualitatively different ways of influencing the environment: some interventions act immediately, whereas others induce persistent effects that continue to shape future states long after the decision that init...

📖 Read original article


190. Local Cluster Cardinality Estimation for Adaptive Mean Shift ​

Author: 'Etienne Pepin
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2508.12450v2 Announce Type: replace Abstract: This article presents an adaptive mean shift algorithm in which every parameter used at a point is derived from that point's own distance distribution. The distance distribution from a point to all others is used to estimate the cardinality of the ...

📖 Read original article


191. Patch-based Memory Gate Model in Time Series Foundation Model ​

Author: Samuel Yoon, Jongwon Kim, Juyoung Ha, Young Myoung Ko
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.18751v4 Announce Type: replace Abstract: Recently reconstruction-based deep models have been widely used for time series anomaly detection, but as their capacity and generalization capability increase, these models tend to over-generalize, often reconstructing unseen anomalies accurately....

📖 Read original article


192. Variational Mixture of Graph Neural Experts for Alzheimer's Disease Recognition across Frequency Bands in EEG Brain Networks ​

Author: Jun-En Ding, Anna Zilverstand, Shihao Yang, Albert Chih-Chieh Yang, Feng Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.11917v4 Announce Type: replace Abstract: Dementia disorders such as Alzheimer's disease (AD) and frontotemporal dementia (FTD) exhibit overlapping electrophysiological signatures in electroencephalography (EEG) that challenge accurate diagnosis. Existing EEG-based methods are limited by f...

📖 Read original article


193. Adaptive Online Learning with LSTM Networks for Energy Price Prediction ​

Author: Salih Salihoglu, Ibrahim Ahmed, Afshin Asadi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.16898v2 Announce Type: replace Abstract: Accurate prediction of electricity prices is crucial for stakeholders in the energy market, particularly for grid operators, energy producers, and consumers. This study focuses on developing a predictive model leveraging Long Short-Term Memory (LST...

📖 Read original article


194. Reliable Inference in Edge-Cloud Model Cascades via Conformal Alignment ​

Author: Jiayi Huang, Sangwoo Park, Nicola Paoletti, Osvaldo Simeone
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, stat.ML

arXiv:2510.17543v3 Announce Type: replace Abstract: Edge intelligence enables low-latency inference via compact on-device models, but assuring reliability remains challenging. We study edge-cloud cascades that must preserve conditional coverage: whenever the edge returns a prediction set, it should ...

📖 Read original article


195. LoRAQuant: Mixed-Precision Quantization of LoRA to Ultra-Low Bits ​

Author: Amir Reza Mirzaei, Yuqiao Wen, Yanshuai Cao, Lili Mou
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.26690v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has become a popular technique for parameter-efficient fine-tuning of large language models (LLMs). In many real-world scenarios, multiple adapters are loaded simultaneously to enable LLM customization for personalized us...

📖 Read original article


196. Accelerating Time Series Foundation Models with Speculative Decoding ​

Author: Pranav Subbaraman, Fang Sun, Jinxi Yu, Yue Yao, Huacong Tang, Xiao Luo, Yizhou Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.18191v2 Announce Type: replace Abstract: Time series forecasting drives operational decisions under tight latency budgets, and autoregressive time series foundation models (TSFMs) increasingly deliver the most accurate forecasts. That accuracy is paid for at inference, since a horizon of ...

📖 Read original article


197. BrowseSafe: Understanding and Preventing Prompt Injection Within AI Browser Agents ​

Author: Kaiyuan Zhang, Mark Tenenholtz, Kyle Polley, Jerry Ma, Denis Yarats, Ninghui Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2511.20597v2 Announce Type: replace Abstract: The integration of artificial intelligence (AI) agents into web browsers introduces security challenges that go beyond traditional web application threat models. Prior work has identified prompt injection as a new attack vector for web agents, yet ...

📖 Read original article


198. A-3PO: Accelerating Asynchronous LLM Training with Staleness-aware Proximal Policy Approximation ​

Author: Xiaocan Li, Shiliang Wu, Zheng Shen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC

arXiv:2512.06547v4 Announce Type: replace Abstract: Decoupled PPO has been a successful reinforcement learning (RL) algorithm to deal with the high data staleness under the asynchronous RL setting. Decoupled loss used in decoupled PPO improves coupled-loss style of algorithms' (e.g., standard PPO, G...

📖 Read original article


199. Grounding Large Language Models as Generalizable Policies in Network Control ​

Author: Duo Wu, Linjia Kang, Zhimin Wang, Fangxin Wang, Wei Zhang, Chongbo Sun, Xuefeng Tao, Wei Yang, Le Zhang, Wenwu Zhu, Peng Cui, Zhi Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.11839v2 Announce Type: replace Abstract: Designing generalizable control policies that operate reliably under changing conditions is essential for robust network services in modern digital infrastructure. Yet network control remains dominated by specialized policies built from handcrafted...

📖 Read original article


200. Federated Learning for the Design of Parametric Insurance Indices under Heterogeneous Renewable Production Losses ​

Author: Fallou Niakh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2601.12178v2 Announce Type: replace Abstract: We propose a federated learning framework for the calibration of parametric insurance indices under heterogeneous renewable energy production losses. Producers locally model their losses using Tweedie generalized linear models and private data, whi...

📖 Read original article


201. Probably Approximately Correct Maximum A Posteriori Inference ​

Author: Matthew Shorvon, Frederik Mallmann-Trenn, David S. Watson
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.16083v2 Announce Type: replace Abstract: Computing the conditional mode of a distribution, better known as the maximum a posteriori (MAP) assignment, is a fundamental task in probabilistic inference. However, MAP is generally intractable, and remains hard even under many common structural...

📖 Read original article


202. Superposition Without Interference? Towards Isolated Interventions via Almost Orthogonal Features in Language Models ​

Author: Moritz Miller, Florent Draye, Bernhard Sch"olkopf
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2602.04718v5 Announce Type: replace Abstract: A central premise in mechanistic interpretability is that meaningful concepts in language models are represented by linear features in activation space. For such features to support reliable interventions, manipulating one feature should not substa...

📖 Read original article


203. A Theoretical Framework for Modular Learning of Robust Generative Models ​

Author: Corinna Cortes, Mehryar Mohri, Yutao Zhong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2602.17554v3 Announce Type: replace Abstract: Training large-scale generative models is resource-intensive and relies heavily on heuristic dataset weighting. We address two fundamental questions: Can we train Large Language Models (LLMs) modularly, combining small, domain-specific experts to m...

📖 Read original article


204. Representation Finetuning for Continual Learning ​

Author: Haihua Luo, Xuming Ran, Tommi K"arkk"ainen, Huiyan Xue, Zhonghua Chen, Qi Xu, Fengyu Cong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.11201v3 Announce Type: replace Abstract: The world is inherently dynamic, and continual learning aims to enable models to adapt to ever-evolving data streams. While pre-trained models have shown powerful performance in continual learning, they still require finetuning to adapt effectively...

📖 Read original article


205. Stochastic Dimension Zeroth-Order Estimator: Stable and Memory-Efficient Training of PINNs ​

Author: Zhangyong Liang, Huanhuan Gao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.24002v4 Announce Type: replace Abstract: Physics-Informed Neural Networks (PINNs) for high-dimensional and high-order partial differential equations (PDEs) are primarily constrained by the $\mathcal{O}(d^k)$ spatial derivative complexity and the $\mathcal{O}(P)$ memory overhead of backpro...

📖 Read original article


206. Weaves, Wires, and Morphisms: Formalizing and Implementing the Algebra of Deep Learning ​

Author: Vincent Abbott, Gioele Zardini
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.CT

arXiv:2604.07242v3 Announce Type: replace Abstract: Despite deep learning models running well-defined mathematical functions, we lack a formal mathematical framework for describing model architectures. Ad-hoc notation, diagrams, and pseudocode poorly handle nonlinear broadcasting and the relationshi...

📖 Read original article


207. Spend Less, Fit Better: Budget-Efficient Scaling Law Fitting via Active Experiment Selection ​

Author: Sijie Li, Shanda Li, Haowei Lin, Weiwei Sun, Ameet Talwalkar, Yiming Yang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.22753v2 Announce Type: replace Abstract: Scaling laws are used to plan multi-million-dollar training runs, but fitting those laws can itself cost millions. In modern large-scale workflows, assembling a sufficiently informative set of pilot experiments is already a major budget-allocation ...

📖 Read original article


208. Optimized Deferral for Imbalanced Settings ​

Author: Corinna Cortes, Anqi Mao, Mehryar Mohri, Yutao Zhong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2604.27723v2 Announce Type: replace Abstract: Learning algorithms can be significantly improved by routing complex or uncertain inputs to specialized experts, balancing accuracy with computational cost. This approach, known as learning to defer, is essential in domains like natural language ge...

📖 Read original article


209. Mind the Gap: Structure-Aware Consistency in Preference Learning ​

Author: Mehryar Mohri, Yutao Zhong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2604.27733v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human intent, whether through explicit reward modeling or direct methods such as DPO, fundamentally relies on minimizing a surrogate loss as a proxy for the true pairwise ranking objective. We prove that t...

📖 Read original article


210. Linear-Core Surrogates: Smooth Loss Functions with Linear Rates for Classification and Structured Prediction ​

Author: Mehryar Mohri, Yutao Zhong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2604.27742v2 Announce Type: replace Abstract: A fundamental dichotomy in the theory of classification sets smoothness against statistical efficiency: smooth surrogate losses such as the logistic loss enable fast $O(1/T)$ optimization but yield slow square-root $H$-consistency bounds, while pie...

📖 Read original article


211. Analytic Bridge Diffusions for Controlled Path Generation ​

Author: Michael Chertkov
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, cs.AI, cs.SY, eess.SY, math.OC

arXiv:2605.02961v2 Announce Type: replace Abstract: Most modern bridge-diffusion methods achieve finite-time transport by specifying an interpolation, Schrodinger-bridge, or stochastic-control objective and then learning the associated score or drift field with a neural network. In contrast, we iden...

📖 Read original article


212. Pretraining large language models with MXFP4 on Native FP4 Hardware ​

Author: Musa Cim, Sarthak Arora, Poovaiah Palangappa, Miro Hodak, Ravi Dwivedula, Meena Arunachalam, Mahmut Taylan Kandemir
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.09825v4 Announce Type: replace Abstract: Why does full-pipeline FP4 training of large language models often diverge, even when forward activations and activation gradients remain stable? We address this question through a controlled study of MXFP4 quantization in transformer training, pro...

📖 Read original article


213. Modeling Spectral Energy Shifts in Spatio-Temporal Graph Anomaly Detection ​

Author: Yilin Liu, Hongchao Zhang, Taylor T. Johnson, Ahmad F. Taha, Meiyi Ma
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.00304v2 Announce Type: replace Abstract: Graph anomaly detection methods aim to distinguish anomalous nodes. While prior methods characterize anomalies through increased variation in the spectral energy distributions, they overlook those that result in decreased variation, i.e., camouflag...

📖 Read original article


214. Planar Symmetric Pattern Generation ​

Author: Ning Lin, Luxi Chen, Huaguan Chen, Jiacheng Cen, Chongxuan Li, Wenbing Huang, Hao Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.02073v2 Announce Type: replace Abstract: Generating objects with specific symmetries is essential in various real-world scenarios. However, adapting existing 2D continuous representations to enforce planar group symmetry remains a challenge, as the transformation of non-reflective group e...

📖 Read original article


215. Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement Learning ​

Author: Xuekang Wang, Zhuoyuan Hao, Shuo Hou, Hao Peng, Juanzi Li, Xiaozhi Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2606.04923v2 Announce Type: replace Abstract: Rubric-based reinforcement learning (RL) uses an LLM-as-a-Judge (LaaJ) to score model outputs according to rubrics as rewards. However, policy models may exploit latent biases in the judge, leading to reward hacking and ineffective or unsafe traini...

📖 Read original article


216. Bootstrap Theory of Representational Emergence: Explanatory Insufficiency as a Driver of Representation Learning and World Models ​

Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.07303v3 Announce Type: replace Abstract: Representation learning is central to modern machine learning, yet most research focuses on optimizing representations after a representational framework has been selected. Less attention is given to when a new representational level becomes necess...

📖 Read original article


217. Detecting Explanatory Insufficiency in Learned Representations: A Framework for Representational Vigilance ​

Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.13172v3 Announce Type: replace Abstract: Learned representations are central to modern machine learning and are typically evaluated through predictive performance, robustness, uncertainty estimation, and generalization. However, a learned representation may remain operationally successful...

📖 Read original article


218. ATMA: Long-Context Language Modeling via Polar Attention and Gated-Delta Compression Memory ​

Author: Habibullah Akbar
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.25156v4 Announce Type: replace Abstract: Length extrapolation in language models involves competing objectives: retrieval fidelity, long-document likelihood, short-context quality, and inference cost. We present ATMA, a 378M-parameter hybrid recipe that combines Polar Attention with gated...

📖 Read original article


219. Prompt-Driven Exploration ​

Author: Sunshine Jiang, John Marangola, David Zhang, Raghuram Kowdeed, Ruiyang Luo, Nitish Dashora, Richard Li, Pulkit Agrawal, Zhang-Wei Hong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08837v2 Announce Type: replace Abstract: Exploration is essential to RL since a policy cannot improve by repeatedly sampling the behaviors it already prefers. Standard methods inject stochasticity in the action space, but such jitter only yields rollouts close to the original. Escaping a ...

📖 Read original article


220. Quantum Port-Hamiltonian Neural Networks: Learning Conservative and Dissipative Dynamics via Measurement-Induced Nonlinearity ​

Author: Dibakar Sigdel
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12269v3 Announce Type: replace Abstract: We introduce Quantum Port-Hamiltonian Neural Networks (Q-pHNNs), parameterised quantum circuits that learn classical dynamics in a structure-preserving manner. The framework rests on the Isomorphic Hamiltonian Mapping (IHM): the skew-symmetric inte...

📖 Read original article


221. Gauge-Fixing the Forward-Forward Objective: A Whitened Goodness Derived from a Likelihood-Ratio Account ​

Author: Paolo Giannitrapani
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, eess.IV, stat.ML

arXiv:2607.12501v3 Announce Type: replace Abstract: The Forward-Forward algorithm trains each layer locally, so that a scalar goodness - the sum of squared activations - is high on real inputs and low on contrastive ones. Under an explicit generative model this goodness is the sufficient statistic o...

📖 Read original article


222. Reducing Per-Sample Interference in Stochastic Optimization ​

Author: Apostolos Avranas
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.16261v2 Announce Type: replace Abstract: Modern optimizers combine gradients from the current mini-batch with historical optimization state, such as momentum or adaptive moments. While effective, this standard practice can produce parameter updates that actively increase the loss of indiv...

📖 Read original article


223. Agentic Bayesian Optimization through Surrogate-Augmented Autoresearch ​

Author: Paul Brunzema, Louis Tiao, Nhat Le, Kevin De Angeli, Yao Xuan, Djordje Gligorijevic
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.00316v2 Announce Type: replace Abstract: Bayesian optimization (BO) has become the standard tool for sample-efficient optimization and owes its efficiency to uncertainty-aware search driven by generic statistical priors. Richer domain priors can improve BO in principle, but encoding them ...

📖 Read original article


224. Empowering Credit Risk Detection in Weixin Pay with Billion-Scale Deep Graph Learning ​

Author: Xin Liu, Xiyuan Chen, Chenglong Wu, Xuan Zong, Jun Zhou, Dawei Cheng
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.02168v2 Announce Type: replace Abstract: Credit risk detection, particularly mitigating individual fraud, is crucial for maintaining the stability of digital financial ecosystems. Accurately identifying credit fraud among billions of users is critical for minimizing financial losses and s...

📖 Read original article


225. Unscented KalmanNet: Structure-Preserving Deep Learning with Calibrated Posterior Uncertainty under Incomplete Physics and Unknown Noise ​

Author: Minhyeok Ko, Abdollah Shafieezadeh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, eess.SP, math.NA

arXiv:2608.04201v2 Announce Type: replace Abstract: Nonlinear state estimation requires sequentially fusing model-based predictions with noisy measurements. Under imperfect dynamics and unknown, time-varying noise statistics, this fusion can degrade in both accuracy and statistical consistency. Exis...

📖 Read original article


226. Continual Learning in Transition ​

Author: Zhiyan Hou, Dan Zhang, Tao Feng, Liyuan Wang, Wei Li, Xiangzhao Hao, Hongyan An, Junfeng Fang, Haokai Ma, Zhaohui Xu, Xinyu Tang, Haiyun Guo, Jinqiao Wang, Tat-Seng Chua
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.06216v2 Announce Type: replace Abstract: Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g., training strategies, architectural designs, and weight adaptation. However, emerging paradigms are...

📖 Read original article


227. ED-CSP: Crystal Structure Prediction from Electron Diffraction ​

Author: Germain Poloudenny, Ya"el Fr'egier, Arnaud Demorti`ere
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.06448v2 Announce Type: replace Abstract: Recovering a periodic 3D crystal structure from sparse, unindexed electron diffraction (ED) observations is a challenging generative inverse problem. Existing ED-based learning methods mainly predict crystallographic labels, reconstruct structures ...

📖 Read original article


228. Causal State-Space Model for Causal Inference: Estimating Longitudinal Individual Treatment Effects ​

Author: Abisoye Abidakun, Mingjun Zhong, Georgios Leontidis
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.08288v2 Announce Type: replace Abstract: Estimating counterfactual outcomes over time from longitudinal observational data is central to clinical decision support. Existing methods rely on domain confusion -- adversarial training that renders representations invariant to treatment assignm...

📖 Read original article


229. CurveFP: Co-Designing Numerical Representation and Product Arithmetic for Language Models ​

Author: Ye Qiao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10010v2 Announce Type: replace Abstract: Low-precision formats usually optimize scalar fidelity while inheriting conventional product arithmetic. We introduce CurveFP, a block-scaled family that distributes magnitudes across interleaved logarithmic curves. Uniform curve indices make every...

📖 Read original article


230. Do Judges Behave Like Algorithms? ​

Author: Riya Manchanda, Eric Chen, Chloe Zhu, Cynthia Rudin, Brandon Garrett, Songman Kang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10400v2 Announce Type: replace Abstract: What if judges already behave like algorithms? As artificial intelligence and algorithms are deployed in many settings, including the judicial system, many have debated whether judges should be allowed to rely on them. Instead, we ask whether judge...

📖 Read original article


231. Optimize Cheap, Deploy Strong: Cost-Aware Cross-Tier Transfer for Evolutionary Optimization ​

Author: Tal Oved, Roi Pony, Oshri Naparstek, Udi barzelay
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.NE

arXiv:2608.10694v2 Announce Type: replace Abstract: Evolutionary optimization of LLM prompts and agentic programs (e.g., GEPA) is dominated by fitness evaluation: scoring each candidate runs an answering LLM over a validation set, so the evaluator's price tier dictates total search cost. We restruct...

📖 Read original article


232. Soft-Attention Improves Skin Cancer Classification Performance ​

Author: Soumyya Kanti Datta, Seyed Mohammad Abuzar Hashemi, Sargur N. Srihari, Mingchen Gao
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2105.03358v4 Announce Type: replace-cross Abstract: In clinical applications, neural networks must focus on and highlight the most important parts of an input image. Soft-Attention mechanism enables a neural network toachieve this goal. This paper investigates the effectiveness of Soft-Attenti...

📖 Read original article


233. Heterogeneous transfer learning for high-dimensional regression with feature mismatch ​

Author: Jae Ho Chang, Massimiliano Russo, Subhadeep Paul
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2412.18081v3 Announce Type: replace-cross Abstract: We study Heterogeneous Transfer Learning (HTL) for high-dimensional regression with differing feature sets. Such feature mismatch arises when some variables available in a data-rich source domain are unavailable in a data-poor target domain. ...

📖 Read original article


234. A Variational Analysis of Kernel Learning with Learnable Linear Transformations ​

Author: Yang Li, Feng Ruan
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.CA, math.FA, math.OC

arXiv:2502.11665v3 Announce Type: replace-cross Abstract: The classical kernel ridge regression problem aims to find the best fit for the output $Y$ as a function of the input data $X\in \mathbb{R}^d$, with a fixed choice of regularization term imposed by a given choice of a reproducing kernel Hilbe...

📖 Read original article


235. SteeringSafety: Benchmarking Representation Steering in LLMs Across Safety Perspectives ​

Author: Vincent Siu, Nicholas Crispino, David Park, Nathan W. Henry, Zhun Wang, Yang Liu, Dawn Song, Chenguang Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2509.13450v3 Announce Type: replace-cross Abstract: We introduce SteeringSafety, a benchmark for evaluating representation steering methods across nine safety perspectives spanning 18 datasets. While prior work highlights the general capabilities of representation steering, we focus on safety ...

📖 Read original article


236. End-to-end Differentiable Calibration and Reconstruction for Optical Particle Detectors ​

Author: Omar Alterkait, C'esar Jes'us-Valls, Ryo Matsumoto, Patrick de Perio, Kazuhiro Terao
Published: 8/13/2026, 4:00:00 AM
Categories: hep-ex, cs.LG, physics.ins-det

arXiv:2602.24129v3 Announce Type: replace-cross Abstract: Large-scale homogeneous detectors with optical readouts are widely used in particle detection, with Cherenkov and scintillator neutrino detectors as prominent examples. Analyses in experimental physics rely on high-fidelity simulators to tran...

📖 Read original article


237. Post-Training with Policy Gradients: Optimality and the Base Model Barrier ​

Author: Alireza Mousavi-Hosseini, Murat A. Erdogdu
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2603.06957v2 Announce Type: replace-cross Abstract: We study post-training linear autoregressive models with outcome and process rewards. Given a context $\boldsymbol{x}$, the model must predict the response $\boldsymbol{y} \in Y^N$, a sequence of length $N$ that satisfies a $\gamma$ margin co...

📖 Read original article


238. LLM Router: Rethinking Routing with Prefill Activations ​

Author: Tanay Varshney, Annie Surla, Michelle Xu, Gomathy Venkata Krishnan, Maximilian Jeblick, David Austin, Neal Vaidya, Davide Onofrio
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2603.20895v3 Announce Type: replace-cross Abstract: Existing routers rely on semantic query features or handcrafted features, which often fail to capture model-specific failures or intrinsic task difficulty. We instead route using internal LLM activations, specifically the residual stream. Our...

📖 Read original article


239. CGRL: Causal-Guided Representation Learning for Node-Level Out-of-Distribution Generalization ​

Author: Bowen Lu, Lianqiang Yang, Teng Li, Kun Zhang
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2603.24304v3 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) deliver strong performance on graph tasks, but their accuracy drops significantly under out-of-distribution (OOD) scenarios. Under distribution shifts, GNNs often fit environmental noise and spurious correlations ...

📖 Read original article


240. Trust Region Constrained Bayesian Optimization with Penalized Constraint Handling ​

Author: Raju Chowdhury, Tanmay Sen, Biswabrata Pradhan
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2603.24567v2 Announce Type: replace-cross Abstract: Constrained optimization in high-dimensional black-box settings is difficult due to expensive evaluations, the lack of gradient information, and complex feasibility regions. In this work, we propose a Bayesian optimization method that combine...

📖 Read original article


241. Diffusion-Guided Cooperative Policy Learning for Target Tracking Based on Underwater Mobile Agent Networks ​

Author: Jiaao Ma, Chuan Lin, Guangjie Han, Shengchao Zhu, Zhenyu Wang, Chen An
Published: 8/13/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2603.29426v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) provides a promising solution for cooperative target tracking in networks of autonomous underwater vehicles (AUVs). However, existing methods still face three major challenges: 1) policy non-stationar...

📖 Read original article


242. On Data-Driven Koopman Representations of Nonlinear Delay Differential Equations ​

Author: Santosh Mohan Rajkumar, Dibyasri Barman, Kumar Vikram Singh, Debdipta Goswami
Published: 8/13/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.DS

arXiv:2604.03086v2 Announce Type: replace-cross Abstract: This work establishes a rigorous bridge between infinite-dimensional delay dynamics and finite-dimensional Koopman learning, with explicit and interpretable error guarantees. While Koopman analysis is well-developed for ordinary differential ...

📖 Read original article


243. ReCodeAgent: A Multi-agent Workflow for Language-Agnostic Translation and Validation of Large-Scale Repositories ​

Author: Ali Reza Ibrahimzada, Brandon Paulsen, Daniel Kroening, Reyhaneh Jabbarvand
Published: 8/13/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2604.07341v3 Announce Type: replace-cross Abstract: Most repository-level code translation and validation techniques have been evaluated on a single source-target programming language (PL) pair, owing to the complex engineering effort required to adapt new PL pairs. Programming agents can enab...

📖 Read original article


244. TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning ​

Author: Matthew M. Hong, Jesse Zhang, Anusha Nagabandi, Abhishek Gupta
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2605.12236v2 Announce Type: replace-cross Abstract: Fine-tuning pre-trained robot policies with reinforcement learning (RL) often inherits the bottlenecks introduced by pre-training with behavioral cloning (BC), which produces narrow action distributions that lack the coverage necessary for do...

📖 Read original article


245. Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models ​

Author: Qinwu Xu, Yifan Jiang, Haoyu Ren
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2605.16409v3 Announce Type: replace-cross Abstract: Optical character recognition (OCR) and multilingual scene-text understanding remain challenging for multimodal large language models (MLLMs), particularly in real-world images containing small or degraded text, cluttered layouts, occlusion, ...

📖 Read original article


246. Large language models reorganize representational geometry during in-context learning ​

Author: Hua-Dong Xiong, Li Ji-An, Robert C. Wilson, Kwonjoon Lee, Xue-Xin Wei
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, q-bio.NC

arXiv:2605.28854v3 Announce Type: replace-cross Abstract: Large language models (LLMs) show remarkable flexibility in adapting to novel tasks without parameter updates, a capacity known as in-context learning (ICL). Prior work has sought to understand ICL by studying the circuits, algorithms, and re...

📖 Read original article


247. Moxia: A Trust-First Neuro-Symbolic Execution Architecture for Self-Explaining Mathematical Reasoning ​

Author: Alessio Bruno
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2606.00671v3 Announce Type: replace-cross Abstract: We present Moxia (formerly AXIOM), a trust-first neuro-symbolic architecture for self-explaining mathematical reasoning over natural-language input. Its language model is strictly a canonicalizer: it rewrites informal problem text into a narr...

📖 Read original article


248. Ranking vs. Assignment: The Metric Mismatch in Multi-View Object Association ​

Author: Matvei Shelukhan, Timur Mamedov, Aleksandr Chukhrov, Karina Kvanchiani
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2606.02022v2 Announce Type: replace-cross Abstract: Multi-view object association is an important computer vision problem that underlies many multi-camera perception tasks. While this task is naturally formulated as a constrained one-to-one matching problem, recent works heavily rely on pairwi...

📖 Read original article


249. RedditPersona: A Modular Framework for Community-Conditioned LLM Adaptation from Reddit ​

Author: Amirhossein Ghaffari, Ali Goodarzi, Huong Nguyen, Simo Hosio, Lauri Lov'en, Ekaterina Gilman
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.SI

arXiv:2606.06027v2 Announce Type: replace-cross Abstract: Community-conditioned language model adaptation needs choices about data collection, community definition, and evaluation that are currently made independently in each study, making it hard to compare assumptions or reuse artifacts. We presen...

📖 Read original article


250. FACTR 2: Learning External Force Sensing for Commodity Robot Arms Improves Policy Learning ​

Author: Steven Oh, Jason Jingzhou Liu, Tony Tao, Philip Han, Kenneth Shaw, Satoshi Funabashi, Ruslan Salakhutdinov, Deepak Pathak
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, cs.SY, eess.SY

arXiv:2606.12406v2 Announce Type: replace-cross Abstract: Contact-rich manipulation requires force sensitivity, but many robot arms lack dedicated force sensors due to their high cost. We present Neural External Torque Estimation (NEXT), a data-driven method that estimates external joint torques wit...

📖 Read original article


251. MMLA: How Memory Lets the Past Shape the Future ​

Author: Junyi Zou, Avrova Donz
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2606.28876v3 Announce Type: replace-cross Abstract: Proposal. Long context can replay history, but it does not decide which completed observations deserve authority. MMLA formalizes a bounded resident memory between transient context and slow weight updates. A completed local segment is eventi...

📖 Read original article


252. Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop ​

Author: Chenmu Zhang, Boris I. Yakobson
Published: 8/13/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.AI, cs.LG

arXiv:2606.29717v2 Announce Type: replace-cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has produced standard public benchmarks and many published machine-learning models for the task (D...

📖 Read original article


253. WBMM: Windowed Batch Matrix Multiplication for Efficient Large Receptive Field Convolution ​

Author: Wan Song, Wei Zhou, Rui Wang, Jun Yu, Toru Kurihara, Jiajia Xu, Shu Zhan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.02097v2 Announce Type: replace-cross Abstract: Large kernel depthwise convolutions achieve strong performance but suffer from significant degradation as kernel size grows due to irregular memory access from gather-based computation; while Large Kernel Acceleration (LKA) helps on small fea...

📖 Read original article


254. OrderMoE: An expert similarity driven distributed edge MoE inference ​

Author: Xin Yuan, Ning Li, Quan Chen, Wenchao Xu, Song Guo
Published: 8/13/2026, 4:00:00 AM
Categories: cs.NI, cs.DC, cs.LG

arXiv:2607.17154v2 Announce Type: replace-cross Abstract: Although mixture-of-experts, MoE, models have been increasingly adopted to scale large language models with moderate computation cost, it remains challenging to deploy MoE inference over resource-constrained and bandwidth-limited edge infrast...

📖 Read original article


255. PathRIR: Physics-Guided Acoustic Path Selection and Late-Tail Compensation for Fast Room Impulse Response Simulation ​

Author: Shaoheng Xu, Chunyi Sun, Jihui Zhang, Amy Bastine, Prasanga N. Samarasinghe, Thushara D. Abhayapala
Published: 8/13/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, cs.SD

arXiv:2607.23293v2 Announce Type: replace-cross Abstract: Image-source-method (ISM)-based room impulse response (RIR) simulation is a useful and physically interpretable tool for acoustic scene modeling, but full-order ISM becomes computationally expensive as the reflection order and room complexity...

📖 Read original article


256. DREvo: Distilling Recalibrated Historical Experience for Harness Self-Evolution ​

Author: Hanghui Guo, Weijie Shi, Zhangze Chen, Shengxiang Xu, Yishu Wang, Yimei Zhang, Wangze Ni, Jia Zhu, Shimin Di
Published: 8/13/2026, 4:00:00 AM
Categories: cs.MA, cs.LG

arXiv:2607.26722v2 Announce Type: replace-cross Abstract: Harness plays a critical role in large language model agent performance, and building a high-performing harness requires substantial expert effort. Therefore, recent research has increasingly explored harness self-evolution, which iteratively...

📖 Read original article


257. A Constitution-Grid Instrument for Data-Efficient RL Alignment (C-Guard) ​

Author: Xianling Zhang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.00180v4 Announce Type: replace-cross Abstract: Conflicting objectives are general in RL alignment, and training on them data-efficiently is hard. Training a safety guard with RL means optimizing two objectives that conflict: catch real harm, and do not refuse benign prompts. Our finding i...

📖 Read original article


258. Rethinking Medical Landmark Localization with Prototype Learning-based Progressive Offset Correction ​

Author: Jingxian Xu, Yuhao Huang, Rusi Chen, Yanfeng Zhou, Dong Ni
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.09182v2 Announce Type: replace-cross Abstract: Accurate landmark localization in medical images is a fundamental step for quantitative clinical measurement and downstream analysis. Existing localization methods have advanced, among which multi-stage refinement is a superior solution. Alth...

📖 Read original article


259. Logit-Boundary Geometric Belief Interfaces and Sparse Sheaf-Enclave Protocols: A Self-Contained Substrate for Secure Network Electronic Health Record (EHR) Interoperability ​

Author: Alvin Spivey, Yu Huang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG

arXiv:2608.10300v2 Announce Type: replace-cross Abstract: Electronic health-record interoperability is a boundary problem: legacy systems, generative models, terminology services, identity systems, and human reviewers may each expose rich internal states, while operational exchange requires a narrow...

📖 Read original article


260. When Do Anchor-Based Pointwise LLM Rerankers Help? Retriever Quality, Statistical Scope, and Anchor Design ​

Author: Utshab Kumar Ghosh, Shubham Chatterjee
Published: 8/13/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.10528v2 Announce Type: replace-cross Abstract: Anchor-based pointwise LLM reranking scores each candidate against a shared reference passage to recover cross-document context at pointwise cost. We study when this actually helps, using GCCP/PAGC as a representative method. Our study is rep...

📖 Read original article


261. Beyond Fixed Luminance: Towards Panchromatic and Orthochromatic Image Colorization ​

Author: Swarnim Maheshwari, Syed Imam Ali, Vineeth N. Balasubramanian
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.10798v2 Announce Type: replace-cross Abstract: Most image colorization systems operate in $Lab$ space by predicting chroma ($ab$) while preserving an input-derived luminance channel ($L$). While effective on standard benchmarks, this fixed-luminance design restricts brightness changes and...

📖 Read original article