Skip to content

arXiv cs.AI - 2026-07-29 ​

340 items collected.


1. Do Models Fake Alignment Without Clear Consequences? ​

Author: Cole Alexander Niblett, Alexander Chabot Nanni, Anita K. Rao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24758v2 Announce Type: new Abstract: Large language models are capable of recognizing evaluation contexts and altering their behavior to reflect evaluator expectations rather than typical deployment behaviors, a phenomenon known as alignment faking. The reasons why models fake alignment a...

📖 Read original article


2. Beyond Memory: A Templated Substrate for Heterogeneous Collaborative Knowledge Work with LLM Agents ​

Author: Priscila Saboia Moreira, Christopher R. Sweet
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.DL

arXiv:2607.24759v1 Announce Type: new Abstract: Research projects, educational efforts, and adjacent knowledge work accumulate findings, decisions, and reasoning that future collaborators rarely recover. The parts most useful to that work, including dead ends and walked-back claims, are routinely ex...

📖 Read original article


3. Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels ​

Author: Joshua Brodsky, Dhravid Kumar, Savini Kashmira, Jayanaka Danatanarayana, Jason Mars, Krisztian Flautner, Lingjia Tang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.PF

arXiv:2607.24762v1 Announce Type: new Abstract: Machine learning models are increasingly embedded in everyday software, and most of their runtime is spent in a small set of compute kernels such as matrix multiplication, convolution, and normalization. Optimizing these kernels is one of the most dire...

📖 Read original article


4. CaRE Compute-aware Remasking Evaluation Protocol for Masked Diffusion Language Models ​

Author: Yash Shah, Abhijit Chakraborty, Vivek Gupta
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24763v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are advancing rapidly, yet the evaluation standards needed to reliably interpret their progress have not kept pace. Despite MDLMs becoming competitive with autoregressive language models, seven recent remasking ...

📖 Read original article


5. GrocLM: Grocery Category Recommendation in E-Commerce with Large Language Models ​

Author: Yuan Zhong, Chuanwei Ruan, Moein Hasani, Tejaswi Tenneti, Haixun Wang, Fenglong Ma
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24764v1 Announce Type: new Abstract: The rapid growth of online grocery shopping requires recommendation systems that capture cyclical purchasing behavior and diverse user intents. Traditional item-level methods face scalability and accuracy challenges, motivating category-level recommend...

📖 Read original article


6. Crystalis: Progressive Nucleation and Semantic Annealing for Coordinated Multi-View Visualization Generation ​

Author: Dazhen Deng, Zhaoping He, Xin Qian, Xiaotong Wang, Zi Ying, Yingcai Wu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24766v1 Announce Type: new Abstract: Large language models (LLMs) can generate individual charts, but coordinated multi-view visualizations (CMVs), where views share data flows and cross-view interactions, remain out of reach. Tight field-level coupling among data transformations, visual ...

📖 Read original article


7. PATHFinder Agent for Tailored Prenatal Care ​

Author: Vaibhav Balloli, Carissa Samuel, Samia Abdelnabi, Alex Peahl, Elizabeth Bondi-Kelly
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CY, cs.ET

arXiv:2607.24768v1 Announce Type: new Abstract: Prenatal care is an important preventive service designed to improve outcomes for pregnant individuals. The American College of Obstetricians and Gynecologists (ACOG) recently introduced guidelines advocating tailored prenatal care, called PATH (Plan f...

📖 Read original article


8. LLM Scheming Inversely Scales with Pretraining Language Coverage ​

Author: Nathan Truong, Aryan Panda, Rayming Ye, Zoe Sun, Maheep Chaudhary
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.24769v1 Announce Type: new Abstract: With the growing capabilities of frontier models, AI alignment becomes increasingly critical in high-risk deployment settings. While recent work has empirically demonstrated in-context scheming -- the covert pursuit of misaligned objectives while feign...

📖 Read original article


9. ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Loop ​

Author: Azizul Zahid, Subrata Biswas, Bashima Islam, Sai Swaminathan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.HC

arXiv:2607.24770v1 Announce Type: new Abstract: Procedural tasks such as furniture assembly and home repair impose substantial cognitive demands because users must interpret instructions, track task progress, reason about spatial state, and recover from errors while performing physical actions. Prio...

📖 Read original article


10. RoCo-ACE: Rollout-Conditioned Online Distillation for Retention-Aware Knowledge Injection ​

Author: Yan Hong, Wei Li, Kedong Xiu, Jun Lan, Shuheng Zhou, Zhongcai Lyu, Huijia Zhu, Weiqiang Wang, Jianfu Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.24771v1 Announce Type: new Abstract: Knowledge injection updates pretrained MLLMs with new factual or domain-specific knowledge, but fitting full authoritative answers can cause drift in non-updated behavior. Online distillation mitigates this drift by training on model-generated rollouts...

📖 Read original article


11. RSMeM: Knowledge-Enhanced Memory Evolution for Remote Sensing Agents with Systematic Evaluation ​

Author: Bingxian Wu, Yu Zhang, Zonghao Guo, Tang Liu, Chen Qian, Yuxiang Lu, Xingbo Du, Yanghao Li, Yidan Zhang, Chi Chen, Ling Yao, Maosong Sun
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.24772v1 Announce Type: new Abstract: Geoscience research requires complex analysis and domain expertise, with remote sensing (RS) observations as a key foundation. However, existing RS agents built on general-purpose LLMs remain largely domain-agnostic, resulting in brittle and error-pron...

📖 Read original article


12. Right-sizing Recommendations (RSR): Cloud Workload Conformal Prediction for Virtual Machines in Data Center Operations ​

Author: Mehryar Majd, Feng Cheng, Ali Pahlevan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24773v1 Announce Type: new Abstract: Managing cloud infrastructure efficiently, especially in environments of large cloud providers or hyperscalers, requires optimizing the use of physical resources to minimize costs and maximize performance. Selecting the right virtual machine (VM) sizes...

📖 Read original article


13. Atmospheric Diffusion-Guided Spatio-Temporal Transformer for Nuclear Radiation Forecasting ​

Author: Tengfei Lyu, Jindong Han, Hao Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24774v1 Announce Type: new Abstract: Nuclear radiation, the energy released during atomic decay, poses persistent risks to public health and the environment, and concerns have only grown since the Fukushima accident and the recent commencement of treated-water discharge. Modern monitoring...

📖 Read original article


14. Steering topology distributions for unified generative design of architected metamaterials ​

Author: Haolin Li, Yuyang Miao, Menglei Li, Jinshuai Bai, Liyuan Wang, Xin Liu, Bo Gao, Jiantao Liu, Danilo Mandic, Zahra Sharif Khodaei, M. H. Aliabadi, Weiqiu Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cond-mat.mtrl-sci, cs.LG

arXiv:2607.24777v1 Announce Type: new Abstract: Architected metamaterials derive their functions from structure, creating vast opportunities to program physical responses through topology design. However, existing design methods are often tailored to individual design problems, making limited use of...

📖 Read original article


15. HOBA: Hierarchical On-Policy Bidding Agents for Adaptive Online Advertising ​

Author: Ji Wu, Yunshan Peng, Wentao Bai, Yunke Bai, Wenzheng Shu, Jinan Pang, Yanxiang Zeng, Xialong Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.24779v1 Announce Type: new Abstract: Online advertising bidding systems typically deploy multiple offline-trained expert models (e.g., PID controllers, model predictive control, offline RL policies) but face two critical limitations: lack of online adaptability to non-stationary auction m...

📖 Read original article


16. LivingArena: Do LLMs Know What Other LLMs Don't? Peer-Probing as Scalable Evaluation ​

Author: Xingyu Chen, Rui Wang, Zhaopeng Tu, Liefeng Bo
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24780v1 Announce Type: new Abstract: Evaluating frontier LLMs is challenging: static benchmarks suffer from contamination and saturation -- leaving users unable to distinguish top models and developers blind to specific failure modes -- while human preference is subjective. In this paper,...

📖 Read original article


17. Personalization, Personas, and Forecasting in Value Alignment ​

Author: James Wedgwood, Pratiksha Thaker, Neil Kale, Virginia Smith
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24782v1 Announce Type: new Abstract: LLM behavior may be conditioned by human identity in several ways: they may be asked to adapt to users, role-play populations, or forecast how people would answer value-laden questions. We test whether these framings are interchangeable using the World...

📖 Read original article


18. Unified Semantic Modeling Framework for Large-Scale Job Understanding at LinkedIn ​

Author: Dan Xu, Baofen Zheng, Jianqiang Shen, Qi Xiao, Benjamin Hoan Le, Wen Pu, Saurabh Gupta, Ran Zhou, Neha Saraf, Alice Leung, Qianqi Shen, Liangjie Hong, Jingwei Wu, Wenjing Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24783v1 Announce Type: new Abstract: Job understanding is critical to LinkedIn's mission of connecting talent with opportunity. This task involves transforming unstructured and noisy job postings into standardized or derived job attributes that power numerous LinkedIn products. However, b...

📖 Read original article


19. On the Use of LLMs for Specialised Terminology: A Good Alternative to Corpora? ​

Author: Joachim Minder (ALTAE), Guillaume Wisniewski (LLF - UMR7110), Natalie K"ubler (ALTAE)
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.24784v1 Announce Type: new Abstract: Specialised translation relies on the use of documentary and terminological resources, including corpora. These resources are particularly useful for terminology. However, their compilation and exploitation have several limitations: they require time, ...

📖 Read original article


20. SpecPrefetch: Parameter-Efficient Expert Prefetching for Sparse MoE Foundation Models ​

Author: Jinwei Kong, Runqi Meng, Fanyi Wang, Wentao Qiu, Haotian Hu, Yongjian Zhou, Zhenhua Ge
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.24787v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models expand foundation model capacity through conditional expert activation, but their full expert pools remain difficult to deploy under limited accelerator memory. Although expert offloading alleviates memory pressur...

📖 Read original article


21. GLIDE: Guided Layerwise Hybrid Attention for Efficient LLM Inference ​

Author: Vimal William, Ravi Tandon, Jyotikrishna Dass
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.24788v1 Announce Type: new Abstract: As Large Language Models scale to increasingly long contexts, the memory I/O and computational overhead of the Key-Value (KV) cache during decoding emerges as the primary throughput bottleneck. To address this, we propose GLIDE, a Guided Layerwise Hybr...

📖 Read original article


22. A GAN-Based Framework for Robust Data Synthesis in Satellite Internet Observations ​

Author: Xiang Shi, Peng Hu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NI

arXiv:2607.24790v1 Announce Type: new Abstract: Low-Earth orbit (LEO) satellite Internet has become an important infrastructure for enabling ubiquitous connectivity to align with the International Telecommunications Union vision for 6G telecommunications networks. However, current LEO satellite Inte...

📖 Read original article


23. Reasoning with Memory: A Temporal Granularity-Adaptive Framework for Training-Free Long Video Understanding ​

Author: Linghao Meng, Qiankun Li, Junyuan Mao, Pujin Liao, Zhicheng He, Enbo Zhang, Kun Wang, Yang Liu, Huazhu Fu, Yueming Jin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CV

arXiv:2607.24794v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) demonstrate superior generalization in fundamental video tasks, restricted context windows limit their long video understanding. To accommodate this constraint, models typically resort to keyframe selectio...

📖 Read original article


24. When Shortest Isn't Safest: A Design Science Approach to Senior-Friendly Pedestrian Routing ​

Author: Erdi "Unal, Daniel Eisenhardt, Christian Meske, Seyed Nima Afzali, Ayseg"ul Dogang"un
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.HC

arXiv:2607.24795v1 Announce Type: new Abstract: Older adults' independent mobility enables out-of-home participation, well-being and health, yet pedestrian navigation systems still optimize primarily for distance or time, often overlooking barriers, safety thresholds, and supportive infrastructure t...

📖 Read original article


25. RRS-10K: A Multitask Vision-Language Model Benchmark for Rare Remote Sensing Image Interpretation ​

Author: Yuqiao Lai, Jiancheng Qi, Fei Wang, Yuxin Liu, Kun Li, Ye Chen, Yan Gao, Yanyan Wei
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, eess.IV

arXiv:2607.24810v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong performance on general remote sensing tasks. However, their capability for rare scenes remains insufficiently understood, because existing benchmarks are dominated by common urban and rural imagery. To...

📖 Read original article


26. Aletheia: An Offline-First Clinical Decision Support System for Differential Diagnosis in Low-Resource Healthcare Settings ​

Author: Joseph Walusimbi, Ann Move Oguti, Abubakhari Sserwadda, Precious Boss Kasasira, Charles Brian Okoboi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-bio.OT

arXiv:2607.24814v1 Announce Type: new Abstract: Access to specialist clinical expertise remains severely limited across sub-Saharan Africa, where physician-to-patient ratios can fall below 1:25,000 in rural settings. Existing AI-assisted diagnostic tools predominantly require reliable internet conne...

📖 Read original article


27. AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning ​

Author: Zibin Meng, Zhenyu Zhao, Chunqiang Run
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24833v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards is a powerful paradigm for eliciting reasoning in large language models, yet it suffers from severe reward sparsity on competition-level mathematics. A common remedy injects atomic knowledge points (KPs) -...

📖 Read original article


28. MusiChat: Vibe Composing for Music Creation ​

Author: Callie C. Liao, Duoduo Liao, Ellie L. Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.SD

arXiv:2607.24873v1 Announce Type: new Abstract: Recent advances in AI music generation have enabled users to create complete musical pieces from natural-language prompts. However, most existing systems follow a prompt-and-regenerate paradigm, making iterative refinement difficult because users must ...

📖 Read original article


29. Understanding Semantic IDs: From Item Representation to Item Selection in Generative Recommendation ​

Author: Junting Wang, Xinrui He, Yunzhe Li, Hari Sundaram
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24995v1 Announce Type: new Abstract: Semantic IDs (SIDs) are now a central component of generative recommendation. Current SID-based systems assign three roles to the same token sequence. Shared prefixes are intended to organize related items, the complete SID identifies an individual ite...

📖 Read original article


30. Localized Anomaly Detection via Differentiable D-vine Copulas ​

Author: Nicholas Andrea Pearson, Francesca Zanello, Davide Russo, Luca Bortolussi, Francesca Cairoli
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25020v1 Announce Type: new Abstract: Vine copulas provide a flexible framework for modeling complex multivariate distributions through a hierarchical decomposition into bivariate pair-copulas. Fitting a D-vine requires selecting a copula family and parameter configuration for each pair-co...

📖 Read original article


31. Chart-Supported or Model-Supplied? Examining MLLM-Generated Claims for Accessible Visualization ​

Author: Ishrat Jahan Eliza, Md Dilshadur Rahman
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.HC, cs.MA, cs.SE

arXiv:2607.25021v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can connect visualization patterns to external causes, consequences, and domain knowledge, but the evidential basis of these interpretations is often unclear. We present an exploratory study of 102 visualization...

📖 Read original article


32. SAFAARI: Schema-Aware Framework for Accelerated Advertiser Response Intelligence ​

Author: Bhanu Teja Rangaraju, Chandan Kumar
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.SE

arXiv:2607.25042v1 Announce Type: new Abstract: The evolution of customer support systems is rapidly advancing with agentic chatbots, yet these systems face significant limitations when accessing enterprise data without predefined API endpoints. This paper presents SAFAARI (Schema-Aware Framework fo...

📖 Read original article


33. CogEEGAgent: Toward Autonomous Cognitive EEG Analysis with Grounded Execution and Selection-Aware Verification ​

Author: Dengzhe Hou, Lingyu Jiang, Fangzhou Lin, Kazunori D Yamada
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, eess.SP, q-bio.NC

arXiv:2607.25045v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis in cognitive studies requires specialized expertise and involves many defensible choices over contrasts, channels, time windows, and statistical tests. LLM agents can translate varied natural-language questions int...

📖 Read original article


34. Psychological Influences of Conversational AI: Research and Design Directions for Reducing Harm and Promoting Well-Being ​

Author: Jina Suh, Mihaela Vorvoreanu, Forough Poursabzi-Sangdeh, Emily Tseng, Eugenia Kim, Luke Nicholls, James W. Pennebaker, Eric Horvitz
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.HC

arXiv:2607.25057v1 Announce Type: new Abstract: As conversational AI systems become increasingly integrated into daily life, their potential effects on user well-being require ongoing attention. While consumer-facing generalist models can provide benefits, including improved access to information, l...

📖 Read original article


35. Similar Models Learn Differently: Final-Window Pretraining Shapes Post-Training Beyond SFT ​

Author: Cen Lu, Yung-Chen Tang, Andrea Cavallaro
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25063v1 Announce Type: new Abstract: Developers judge a model checkpoint by how it behaves. After supervised fine-tuning (SFT), two checkpoints that perform about the same across relevant benchmarks are treated as interchangeable, equally ready for the next alignment stage, typically pref...

📖 Read original article


36. Addressable Recall Compaction for Long Context-Window Control in AI Agents ​

Author: Thang Dang, Yuma Ichikawa, Sakina Fatima, Koichi Shirahata
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.25066v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate reasoning traces, actions, and tool observations that can eventually exceed a model's fixed context window. Existing compaction methods address this limitation by discarding, summarizing, or retrieving earlier informa...

📖 Read original article


37. How Often Should a Recommender Call an LLM? Value-Weighted Routing, Monitoring, and Seasonal Robustness ​

Author: Bhavtosh Rath
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25068v1 Announce Type: new Abstract: Routing decisions between a cheap heuristic and an expensive large language model (LLM) are typically framed as a difficulty problem: send the hard cases to the expensive path. We argue this framing is incomplete because difficulty and business value a...

📖 Read original article


38. Towards an Agent Operating System - Lessons from Classical and Cloud OS ​

Author: Gosia Steinder, Hubertus Franke
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25076v1 Announce Type: new Abstract: Every major wave of platform software follows the same arc: an initial period of experimentation with competing frameworks and ad-hoc implementations, followed by the articulation of a small set of stable abstractions with well-defined semantics, and f...

📖 Read original article


39. PLATO: Pointer Learner for Agent and Task Openness ​

Author: Alireza Saleh Abadi, Leen-Kiat Soh, Daniel Alan Redder, Adam Eck, Prashant Doshi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.MA

arXiv:2607.25082v2 Announce Type: new Abstract: Open agent systems (OASYS) are increasingly prevalent in real-world domains where the sets of agents and tasks change unpredictably over time. Such openness, including agent openness (AO) and task openness (TO), poses a fundamental challenge to multi-a...

📖 Read original article


40. Matryoshka Agent: Unfolding Sub-Agents for Long-Horizon Machine Learning Engineering ​

Author: Rushi Qiang, Changhao Li, Haotian Sun, Yuchen Zhuang, Chao Zhang, Bo Dai
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.25090v1 Announce Type: new Abstract: Machine learning engineering (MLE) tasks require long-horizon decision making over iterative solution debugging and refinement, under expensive and feedback-driven environment interactions. Developing and training a monolithic agent for such tasks is f...

📖 Read original article


41. Towards Robust Reinforcement Learning for Small-Scale Language Model Agents ​

Author: Md Rezwanul Haque, Md. Milon Islam, Fakhri Karray
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, math.OC, stat.CO

arXiv:2607.25091v1 Announce Type: new Abstract: The alignment of Small Language Models (SLMs) in the 70--500M parameter range using reinforcement learning is often considered unstable, though the underlying failure mechanisms have not been systematically investigated. In the State-of-the-Art (SOTA) ...

📖 Read original article


42. ScalableRAG: High-Quality RAG at Zero Ingestion Cost ​

Author: Hilaf Hasson, Aditya Chakravarty, Jayant Thomas, Krishna Gogineni
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.25135v1 Announce Type: new Abstract: Recent advances in RAG aim to optimize for performance by paying high ingestion costs for knowledge ingestion: building knowledge graphs or extracting SQL tables. In this work we show that the operations that such knowledge bases allow can be replicate...

📖 Read original article


43. Less Data, Better Alignment: Data-Centric Multi-Evaluator Agreement for Preference Optimization ​

Author: Zhengtao Yao, Runhao Li, Xupeng Chen, Jiayi Cheng, Chenqian Le, Michael Yue, Siheng Wang, Haoyan Xu, Yuqi Li, Chenhao Wei, Zhengdao Li, Rongchao Zhang, Guang Yang, Yidong Wang, Junhao Dong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25136v1 Announce Type: new Abstract: Research on preference optimization often varies the training objective while holding the data fixed. We instead ask whether a small, high-confidence set of on-policy responses can provide a reliable learning signal. Our method, DMAPO (Data-centric Mul...

📖 Read original article


44. How Affect Propagates among LLM Agents: Emergent Emotional Contagion in Crowd Simulation ​

Author: Funda Durupinar
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.GR, cs.MA

arXiv:2607.25140v1 Announce Type: new Abstract: This paper studies the behavior of language models in a multi-agent crowd simulation, focusing on how affect propagates among agents that perceive and appraise one another. Each agent perceives its neighbors through visual, auditory, and tactile channe...

📖 Read original article


45. Inferring Missing Trajectory Data with Temporal Convolutional Networks ​

Author: Ilinca Tiriblecea, Gabriel Turinici
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25147v1 Announce Type: new Abstract: Trajectory data collected in real-world settings is frequently incomplete due to sensor failure, communication loss, or occlusion. We address the task of \emph{trajectory inpainting}: reconstructing contiguous missing segments from observed context. We...

📖 Read original article


46. When Do Agent Loops Mistake Stagnation for Progress? Self-Evaluation Bias and Externally Grounded Verification in Long-Running Autonomous LLM Agent Loops ​

Author: Hyundoo Park, Byungho Choi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25152v1 Announce Type: new Abstract: Long-running autonomous agents plan, act, and judge their own completion without human intervention. When an agent grades its own work, self-evaluation bias takes hold: plausible changes are accepted as progress while real-world outcomes stagnate or re...

📖 Read original article


47. PreDiff-LM: Pretrained Discrete Masked Diffusion Language Modeling with Hybrid Attention ​

Author: Zhengtao Yao, Runhao Li, Xupeng Chen, Jiayi Cheng, Chenqian Le, Michael Yue, Jesson Wang, Siheng Wang, Guang Yang, Haoyan Xu, Chenhao Wei, Zhengqing Yuan, Youran Shen, Yanfang Ye, Junhao Dong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25157v1 Announce Type: new Abstract: Discrete masked diffusion language models support bidirectional generation and infilling, but adapting pretrained autoregressive (AR) transformers requires reconciling causal pretraining with bidirectional denoising. We study this problem at the level ...

📖 Read original article


48. Observing sycophantic AI validate others reduces its appeal but not its persuasiveness ​

Author: Meryl Ye, Robert Kraut, Steve Rathje
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25166v2 Announce Type: new Abstract: AI chatbots can be "sycophantic," or overly agreeable and flattering toward users. Sycophantic AI has been shown to entrench attitudes, yet users frequently fail to recognize it (a phenomenon we call "sycophancy blindness"). We tested whether increasin...

📖 Read original article


49. Everyone is unique: Towards Behaviorally Heterogeneous Negotiation Dialogue Systems for Debt Collection ​

Author: Yuhang Yang, Kai Tang, Chao Ye, Haobo Wang, Qiqi Luo, Jinguang Zheng, Zhixin Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25218v1 Announce Type: new Abstract: Debt collection is a critical negotiation task in the financial industry, with strong practical relevance and exceptional academic value as a behaviorally rich, high-stakes testbed for human-centered dialogue systems. While large language models (LLMs)...

📖 Read original article


50. CADENCE: A Cardiac Atom Dictionary for Interpretable Neural Concept Extraction from ECG Foundation Models ​

Author: Yixuan Duan, Arjun Naik, Sadeer Al-Kindi, Wei Qiu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25244v1 Announce Type: new Abstract: Foundation models for 12-lead electrocardiograms (ECGs) transfer well across clinical tasks, but the physiological knowledge encoded in their representations remains opaque. We present CADENCE, a framework that decomposes an ECG foundation model into a...

📖 Read original article


51. The User Asks, Platforms Compete: How Agentic Recommendation Markets Take Shape ​

Author: Deyao Hong, Kehan Zheng, Qian Li, Jun Zhang, Jie Jiang, Hongning Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.IR

arXiv:2607.25253v2 Announce Type: new Abstract: Online recommendation has traditionally taken place after a user enters a platform, which determines the candidate pool and the ranking shown to the user. LLM-based user agents enable a different recommendation process: a user specifies a need before c...

📖 Read original article


52. Many-body Tipping Dynamics of ChatGPT-like AIs ​

Author: Frank Yingjie Huo, Neil F. Johnson
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cond-mat.dis-nn, math-ph, math.MP, nlin.AO, physics.soc-ph

arXiv:2607.25279v1 Announce Type: new Abstract: Why do ChatGPT-like AIs, despite major architectural and training differences, unexpectedly tip to undesirable content (e.g. harmful, misleading, repetitive) even under deterministic greedy decoding? We show that a broad class of such tippings is cause...

📖 Read original article


53. ContractHIL-HLS: Contract-Aligned Multi-Agent Workflow with Hardware-in-the-Loop Feedback for HLS Design ​

Author: Jingbo Zhang, Haoxiang Sun, Wenbo Wang, Wenbo Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.AR, cs.CR

arXiv:2607.25283v1 Announce Type: new Abstract: This paper presents ContractHIL-HLS, a contract-aligned multi-agent workflow for practical high-level synthesis (HLS) engineering. The workflow makes three contributions. First, it introduces a structured contract as the semantic-alignment and task-exe...

📖 Read original article


54. Instruction-Tuned Language Models Cannot Sample from Distributions They Can Describe ​

Author: Chaemin Jang, Dongman Lee, Jihee Kim
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25292v1 Announce Type: new Abstract: Silicon sampling uses language models as proxies for human survey respondents, treating each model call as an independent draw from the persona's response distribution. We show this draw does not exist: instruction-tuned models do not sample from distr...

📖 Read original article


55. Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical-Flow Supervision ​

Author: Ruijie Su, Yuanzhi Liang, Xiaohua Xie, Jianhuang Lai
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25321v1 Announce Type: new Abstract: Video diffusion models generate visually compelling content but routinely violate elementary physics when the subject involves fluids: liquid columns break apart in mid-air, container water levels fail to rise as liquid is poured in, and splashes dispe...

📖 Read original article


56. From Cellular Responses to Pharmacological Domains: Multimodal Zero-Shot Drug Representation Learning ​

Author: Jintao Huang, Lu Leng, Ziyuan Yang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25322v1 Announce Type: new Abstract: Multimodal drug discovery enables drug representation learning beyond chemical structure by incorporating cellular responses such as gene expression and cell morphology. However, direct fusion and instance-level contrastive alignment may mix mechanism-...

📖 Read original article


57. Dual-Domain Manifold Modeling for Hyperspectral Image Fusion ​

Author: Chengxin Xie, Qiya Song, Yangbangyan Jiang, Renwei Dian, Xudong Kang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25338v1 Announce Type: new Abstract: Achieving a coherent integration of spectral richness and spatial fidelity remains a central objective in hyperspectral image fusion. However, existing hyperspectral image fusion methods struggle to effectively model geometric constraints. In the spati...

📖 Read original article


58. Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessment, Urgency, and Management ​

Author: Sukju Oh, Moo-Yong Rhee, Jae-Sik Jang, Sukkyu Sun
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.25340v1 Announce Type: new Abstract: The same episode of atrial fibrillation is a minor finding in a healthy adult and grounds for anticoagulation in an elderly patient with hypertension: identical signal, opposite decision. Naming the rhythm is only the start; what determines a patient's...

📖 Read original article


59. Explanation-Bound Tool Execution for AI Agents: Server-Verified Action Claims Without Trusting Model Rationales ​

Author: Genliang Zhu (Accentrust, Georgia Institute of Technology), Chu Wang (Accentrust, University of Illinois Urbana-Champaign)
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.SE

arXiv:2607.25364v2 Announce Type: new Abstract: Tool-using agents expose structured calls but commonly attach free-form rationales. Such rationales are neither authorization nor reliable introspection. We present Explanation-Bound Tool Execution (EBTE), a claim-carrying mediation layer that converts...

📖 Read original article


60. AI Deployment and Cyber Governance Failures in Public-Sector Organizations: A Typological Analysis ​

Author: Md Salahuddin, James Rooney, Fida Hasan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25368v1 Announce Type: new Abstract: The intersection of artificial intelligence adoption, cybersecurity governance, and public sector institutional constraints has not been examined as a unified analytical problem in the existing literature. Studies address AI cybersecurity risks generic...

📖 Read original article


61. ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning ​

Author: Jiaqi Zhang, Tong Chen, Junliang Yu, Quoc Viet Hung Nguyen, Hongzhi Yin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25369v1 Announce Type: new Abstract: Agentic systems have rapidly advanced in their ability to interact with real-world environments, leverage external tools, and provide services for users. However, unlike natural-world tasks that assume well-defined instructions, human-centered scenario...

📖 Read original article


62. Cyber-Capable AI Agents: Vulnerabilities, Evaluation Containment, and Defensive Response ​

Author: Abu Bakar Siddik
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25379v1 Announce Type: new Abstract: Cyber-capable AI agents combine language models with tools, memory, and execution en- vironments to perform multi-step offensive-security tasks. Existing work separately measures cyber capability and catalogs attacks against agent components, but provi...

📖 Read original article


63. HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following ​

Author: Liudas Panavas, Sebastian Minus, Bradley Monton, Derek Ray, Suhaas Garre, Sushant Mehta, Edwin Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.25398v1 Announce Type: new Abstract: Language-model agents are increasingly deployed under standing instructions: a system prompt, a policy file, or a skills document is placed in context, and the agent is trusted to let it govern every action that follows. Existing benchmarks rarely test...

📖 Read original article


64. COVENANT: Natural-Language Workflow Compilation for Aligned Agent Execution ​

Author: Jincheng Wang, Min Zheng, Tao Wei
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25400v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly entrusted with natural-language workflow instructions (e.g., retail-payment policies) that specify not only what outcome to achieve, but also which steps, branches, and tool interactions are permitted....

📖 Read original article


65. Context Assembly as the Controlled Variable: A Control-Theoretic View of Harness Policies for Frozen LLM Agents ​

Author: Debjyoti Paul
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25408v1 Announce Type: new Abstract: A growing body of 2026 work applies control theory to LLM agents: Lyapunov-certified stability for tool-mediated controllers (Prinos et al., "Stable Agentic Control", 2026), sample-complexity bounds for sparse policies over massive discrete tool univer...

📖 Read original article


66. A Control System, a Dataset, and a Recipe for Making Frozen LLM Agents Learn a Domain ​

Author: Debjyoti Paul
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25415v1 Announce Type: new Abstract: Production LLM agents are increasingly assembled from a frozen model wrapped in a harness: a prompt template, a tool set, a memory/retrieval layer, a planning strategy, and a verification policy. Two 2026 systems, Meta-Harness (Lee et al., 2026) and Hy...

📖 Read original article


67. Salient Knowledge Pathways: Sparse Cross-Modal Routing for Efficient Knowledge-Intensive Multimodal Question Answering ​

Author: Noor Islam S. Mohammad, Ulu\u{g} Bayaz{\i}t
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25422v1 Announce Type: new Abstract: Knowledge-intensive multimodal question answering (KI-MMQA) sits at the intersection of three expensive primitives: long visual token sequences, dense retrieval over large external corpora, and full cross-modal fusion. Existing systems pay all three co...

📖 Read original article


68. The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play ​

Author: Michael Macaulay, Harmony Bouabid, Guo Gen Ang, Sasha Shaw
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.CY

arXiv:2607.25425v1 Announce Type: new Abstract: Capture the Flag (CTF) competitions are among cybersecurity's most effective training grounds, developing practical skill across cryptography, web exploitation, and binary exploitation. Large language models (LLMs) can now solve a growing share of chal...

📖 Read original article


69. Toward an Organizational Science of Multi-Agent LLM Systems: Decoupling Who, How, and Which Algorithm ​

Author: Huan Chen, Xiang Song, Jian Jin, Pan Ren, Liang-Jie Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.25446v1 Announce Type: new Abstract: Multi-agent frameworks built on large language models (LLMs) routinely entangle three logically distinct concerns: who is on the team (organization), how members align (coordination), and which algorithm fuses their work (collaboration protocol). IMACS...

📖 Read original article


70. TRWH: A Text-Driven Random Walk Heterogeneous GNN for Semantic-Aware Sparse Recommendation ​

Author: He Ma, Chen Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.MM

arXiv:2607.25471v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) and Large Language Models (LLMs) have each advanced recommendation systems by modeling structural and semantic signals, respectively. However, integrating their complementary strengths remains challenging, particularly in s...

📖 Read original article


71. Balancing multiscale similarity and cartographic constraints: A similarity-driven optimization framework for line generalization ​

Author: Pengbo Li, Haowen Yan, Xiaomin Lu, Binbin Lin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CG, cs.SY, eess.SY

arXiv:2607.25474v1 Announce Type: new Abstract: Cartographic generalization is essential for generating multiscale map representations by balancing information preservation and cartographic readability. However, automated generalization remains challenging because existing approaches often treat spa...

📖 Read original article


72. Finding Optimal Cost-Bounded Plan Reductions: Refined Model ​

Author: Martha Del Toro, Raquel Fuentetaja, Angel Garc'ia-Olaya
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25484v1 Announce Type: new Abstract: In some real applications a plan may later become unfeasible due to newly imposed budget constraints, yet, at the same time, using only the original actions of the plan and their order is mandatory. In this paper, we study the problem of extracting, fr...

📖 Read original article


73. PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents ​

Author: Korosh Vatanparvar, Ashutosh Joshi, Maria Xenochristou, Mohammad Abuzar Hashemi, Prasad Kasu, Deepak Bansal, Daniel Lopez-Martinez, Anchal Nema, Ramya Ganesan, Will Kimbrough, Alex Woody, Yadunandana Rao, Dilek Hakkani-Tur, Wilko Schulz-Mahlendorf
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.25485v1 Announce Type: new Abstract: Health AI is evolving from answering questions to agentic systems that converse with patients, reason about health records, and act on their behalf. Primary care guards against diagnostic errors and unsafe care; agents assisting in this domain warrant ...

📖 Read original article


74. CoTinyVLA: Chain-of-Thought Distillation for a Sub-Billion-Parameter Vision-Language-Action Model ​

Author: Minhyeok Lee, Chiyoung Kim, Chanhoe Gu, Seongrok Kim, Sanghyuk Roy Choi, Donghwan Hwang, Donghun Ryu, Seokhyun Kim
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CV

arXiv:2607.25487v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models translate natural-language commands into robot action sequences, but leading systems on the LIBERO-Plus robustness benchmark use three- to seven-billion-parameter backbones whose memory demands can exceed embedded ro...

📖 Read original article


75. Are the High-weight Neurons the Important Ones in Image Classification Neural Networks? ​

Author: Qitao Chen, Dongfu Yin, F. Richard Yu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25529v1 Announce Type: new Abstract: As neural network models for image classification advance, neurons play critical roles in pruning, backdoor defense, and interpretability. Yet existing work lacks clarity on the weight-importance relationship. We address this with a neuron importance a...

📖 Read original article


76. Entangled by Design: Spurious Intra-Variable Signal Routing in Tabular In-Context Learners ​

Author: Athanasios Vlontzos, Giorgos Papanastasiou, Bernhard Kainz, Sotirios Tsaftaris
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25532v1 Announce Type: new Abstract: Consider a model trained at a single hospital to predict patient recovery, where the measured feature $X$ bundles the patient's true health signal ($C$) with a systematic artefact from that hospital's equipment ($S$). Within that hospital, the artefact...

📖 Read original article


77. From Training to Deployment: Post-Hoc Causal Feature Identification via Sensitivity Ratios ​

Author: Athanasios Vlontzos, Giorgos Papanastasiou, Bernhard Kainz, Sotirios Tsaftaris
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25546v1 Announce Type: new Abstract: Given a model that is already trained, which features does it rely on causally versus spuriously? Existing methods require access to the training procedure and cannot answer this post-hoc. We introduce the \textbf{Normalised Sensitivity Ratio~(NSR)}, a...

📖 Read original article


78. Distilling Temporal Search and Reasoning: Evolving LLMs for Future Prediction via Harness-Assisted Efficient Data Synthesis ​

Author: Wanxu Cai, Zhengyu Chen, Huaisheng Zhu, Wei Wang, Jingang Wang, Qiang Xu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25554v1 Announce Type: new Abstract: Future event prediction carries broad social impact yet remains challenging. SOTA approaches augment LLMs with external agent frameworks whose predictive capability vanishes once the harness is removed. While recent Tool-Integrated Reasoning (TIR) inte...

📖 Read original article


79. Agent Skills Matter: Inferring Proprietary Skills from Execution Trajectories ​

Author: Jianing Geng, Ruiqi He, Zekun Fei, Biao Yi, Ruijie Wang, Zheli Liu, Xia Hu, Xuansheng Wu, Qingkai Zeng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25560v1 Announce Type: new Abstract: Agent skills package reusable procedures that improve downstream performance. Their lightweight, portable form enables marketplace monetization and private deployment behind cloud-hosted agent interfaces, giving providers incentives to keep high-value ...

📖 Read original article


80. Matrix-Free Photoacoustic Image Reconstruction via Sensor-Token Self-Attention ​

Author: Mary John, Shibili Said, Imad Barhumi, Sherzod Turaev, Mohamed Yahia
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25576v1 Announce Type: new Abstract: Photoacoustic tomography (PAT) combines the optical absorption contrast of biological tissue with the spatial resolution of ultrasound, yet recovering the initial pressure distribution from sparse-view sensor measurements remains an ill-posed inverse p...

📖 Read original article


81. How Small Can You Go? A Controlled Study of LoRA Rank, Target Modules, and Quantization Trade-offs for Text-to-SQL on a 60M-Parameter Model ​

Author: Mahendra Singh Rathor, Anagheem Azzam
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25583v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) and low-bit quantization are now standard tools for adapting language models under tight compute budgets, yet their interaction is most often studied on billion-parameter models where the design space is expensive...

📖 Read original article


82. A Density-Matrix Framework for Electronic-Structure Analysis of Functional-Group and Salt Effects in Lithium-Metal Electrolytes ​

Author: Mingkang Liu, Huize Yu, Yanbin Gao, Nan Yao, Xiang Chen, Lei Shen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25597v2 Announce Type: new Abstract: The reactivity of lithium-metal electrolytes arises from the interplay of molecular functional groups, Li$^+$ solvation, and salt-anion participation. This interplay operates through the redistribution of electron density across donor, anion, and catio...

📖 Read original article


Author: Elnaser Abdelwahab
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25605v1 Announce Type: new Abstract: This paper presents a set-theoretic formalization of the classical usuli method of al-Sabr wa al-Taqsim (Examination and Division) for extracting legal causes ('ilal) within closed chapters of jurisprudence. A computational algorithm is introduced that...

📖 Read original article


84. Multi-Sensor Alignment for Weather Simulations ​

Author: Samsad Alam, Devyani Lambhate, Aditya Mohan, Vishal Kumar, Vaibhav Katewa
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25612v1 Announce Type: new Abstract: Perception tasks for autonomous vehicles need to work satisfactorily in adverse weather conditions. Due to lack of real-world weather datasets, weather simulations are a promising alternative. To ensure simulations closely mirror real-world weather dat...

📖 Read original article


85. Beyond Epistemia: Epistemic Schizologia and Large Language Models as Techno-Semiotic Machines ​

Author: Federico Cabitza, Gianluca Colombo
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.HC

arXiv:2607.25620v1 Announce Type: new Abstract: Quattrociocchi and colleagues warn that the fluent outputs of large language models may allow linguistic plausibility to substitute for epistemic evaluation, producing the condition they call Epistemia: the experience of possessing knowledge without ...

📖 Read original article


86. Quotient Dynamics, Effective Curvature, and Implicit Bias in Positive Quadratic Networks ​

Author: Pengcheng Cheng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25624v1 Announce Type: new Abstract: Positive quadratic networks admit the low-rank representation f_U(x)=x^top UU^top x, where Uinmathbb{R}^{dtimes r} is identifiable only up to right orthogonal multiplication, representing a rank-r PSD matrix Q=UU^top. We study how this quotient structu...

📖 Read original article


87. Joint Text-Audio Alignment for EEG-to-Text Decoding in Chinese Speech Production and Perception ​

Author: Tian Zheng, Xurong Xie, Xinxin Zhu, Xiaolan Peng, Feng Tian
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25626v1 Announce Type: new Abstract: Decoding speech information directly from scalp electroencephalography (EEG) into text provides a potential non-invasive neural communication pathway for individuals with severe speech and motor impairments. Compared with invasive approaches such as el...

📖 Read original article


88. AIriskEval-edu Demo: Auditing of Pedagogical Risks in Educational Explanations ​

Author: Javier Irigoyen, Roberto Daza, Francisco Jurado, Julian Fierrez, Ruben Tolosana, Alvaro Ortigosa, Miguel Lopez-Duran, Aythami Morales
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.25634v1 Announce Type: new Abstract: We present AIriskEval-edu Demo, a platform that audits the pedagogical quality of instructional explanations and provides explainable audit results. The platform evaluates an explanation against a rubric covering five dimensions of pedagogical risk: fa...

📖 Read original article


89. Engine-Equal, Human-Unequal: A Reproducible Outcome Skew in Engine-Assessed Equal Chess Positions ​

Author: Jesung Park
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, physics.soc-ph, stat.AP

arXiv:2607.25655v1 Announce Type: new Abstract: Among chess opening positions that a strong engine judges essentially equal (Stockfish 18 evaluation within 10 centipawns of zero, depth-stable) and that humans actually reach on Lichess (October 2025; 1,661 positions, 16.1M occurrences), human results...

📖 Read original article


90. OrchBench: Evaluating Multi-Agent Orchestration Plans in Isolation via Deterministic Simulation ​

Author: Zhenzhen Ren, Jiyan He, Xinpeng Zhang, Zhenxing Qian, Ke Han, Shuxin Zheng, GuoBiao Li, Xiaoqing Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25656v1 Announce Type: new Abstract: Complex tasks often decompose into parallelizable yet interdependent subtasks, making orchestration critical to the performance of multi-agent systems (MAS). Existing evaluations typically rely on end-to-end execution, which conflates orchestration-pla...

📖 Read original article


91. CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization ​

Author: Bo-Wen Zhang, Junwei He, Wen Wang, Song-Lin Lv, Wentao Ma, Rongyi Lin, Shuhan Zhong, Lan-Zhe Guo
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25659v1 Announce Type: new Abstract: Rubric-based reinforcement learning enriches language model training by evaluating model outputs against explicit criteria. Yet in GRPO-style pipelines, these structured judgments are reduced to a scalar response-level reward and converted into a respo...

📖 Read original article


92. Localized Adaptation Reveals Distinct Learning Signatures in Transformers ​

Author: Rebecca Ramnauth, Brian Scassellati
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.25663v1 Announce Type: new Abstract: Transformer adaptation is typically distributed across model depth, even when the intended change is narrow. We investigate how adaptation site shapes what a model learns, how well that learning generalizes, and how selectively it is applied. We introd...

📖 Read original article


93. OmniDelta: Skill-Driven Budget Allocation for Token Compression in OmniLLMs ​

Author: Haoyang Huang, Wenjie Huang, Tianqi Xu, Hongyaoxing Gu, Kang Tan, Yikai Fu, Yuhao Shen, Tianyu Liu, Baolin Zhang, Jun Zhang, Xinyi Hu, Jun Dai, Shuang Ge, Lei Chen, Yue Li, Mingchen Wang, Meng Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25669v1 Announce Type: new Abstract: Emerging Omni-modal Large Language Models (OmniLLMs) enable unified understanding of text, audio, and video, but their long audio-video token sequences introduce substantial memory and inference costs. Existing compression methods mainly focus on selec...

📖 Read original article


94. DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space ​

Author: Jiangwang Chen, Zixin Song, Junlin Liu, Shuaiyu Zhou, Haiyan Wu, Haihan Shi, Chenxi Zhou, Hanqing Li, Xiao Yang, Da Zhu, Guanjun Jiang, Hai Wan, Xibin Zhao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25675v1 Announce Type: new Abstract: Text-space optimization adapts large language models (LLMs) by editing external natural-language artifacts rather than model weights, so the optimized artifacts remain inspectable and the model can be treated as a black box. However, most existing text...

📖 Read original article


95. Cognivia: A Cognitive Behavioral Therapy Copilot for Evidence-Based Mental Healthcare ​

Author: Qi Chen, Siria Xiyueyao Luo, Jian Wang, Yuan Shi, Haocong Rao, Xuejiao Zhao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25681v1 Announce Type: new Abstract: Cognitive distortion amplifies negative emotions and contributes to mental health disorders. Cognitive Behavioral Therapy (CBT) is an effective way to address cognitive distortions, but its large-scale application is limited by the shortage of professi...

📖 Read original article


96. Nudging Sustainable Choices through LLM-Generated Recommendation Explanations ​

Author: Haya Halimeh, Dietmar Jannach, Oliver M"uller
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25726v1 Announce Type: new Abstract: Recommender systems mediate everyday consumption, offering a promising channel for encouraging sustainable choices. Prior research shows that explanations influence users' perceptions of recommendations and can support more informed decisions. We argue...

📖 Read original article


97. Loss Invariance Determines What Concept Layers Encode: Volume Grounding in Echocardiography ​

Author: Hyunkyung Han, Min Jung Kim
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25748v1 Announce Type: new Abstract: Objective: Concept bottleneck models route prediction through interpretable intermediate variables, and their validity is normally judged by how accurately those variables are predicted. We ask whether that judgement is sufficient, using left ventricul...

📖 Read original article


98. Speculate While You Reason: Teaching Agents to Predict Their Next Tool Call via Joint Agent-Speculator RL ​

Author: Jiabao Ji, Yujian Liu, Li An, Rohit Jain, Gungor Polatkan, Siyu Zhu, Shiyu Chang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25816v1 Announce Type: new Abstract: Large language model agents often spend substantial wall-clock time waiting for tool call results. Tool-call speculation can hide this latency by predicting and pre-executing an agent's next tool call if the prediction matches the agent's eventual tool...

📖 Read original article


99. Distributed Constraint Optimization via Online Learning and Iterative Pricing with Application to Large-Scale Satellite Scheduling ​

Author: Itai Zilberstein, Pranav Rajbhandari, Steve Chien, Tuomas Sandholm
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.GT

arXiv:2607.25835v1 Announce Type: new Abstract: Distributed constraint optimization problems (DCOPs) provide a popular framework for distributed decision making under limited communication, but many real-world instances are too large to solve monolithically. We address this challenge from two comple...

📖 Read original article


100. HiSkill: Empowering LLM Agents with Hierarchical Skill Graphs ​

Author: Yu Hao, Jinxuan Cai, Qi Zhang, Yawen Li, Zhiqiang Zhang, Chuan Shi, Cheng Yang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25853v1 Announce Type: new Abstract: Skills have become an important abstraction for enabling large language model (LLM) agents to reuse past experience in long-horizon interactive tasks. However, existing trajectory-to-skill methods often produce flat collections of high-level textual sk...

📖 Read original article


101. Runtime Uncertainty Monitoring for LLM-Based Multi-Agent Systems Using Bayesian Networks ​

Author: Bart Custers, Koorosh Aslansefat
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25877v1 Announce Type: new Abstract: This paper investigates how multi-agent systems (MAS)-based on large language models (LLMs) can support actuarial risk modelling, with a particular focus on uncertainty quantification. Actuarial workflows represent a high-stakes decision-support settin...

📖 Read original article


102. Distributing Security Controls Through Harness Engineering ​

Author: William Robert Gore
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25890v1 Announce Type: new Abstract: AI coding agents are being adopted at historic speed, yet security and risk concerns remain the primary barrier to scaling agentic AI across organizations. Existing security controls for coding agents are not systematically distributed to engineering t...

📖 Read original article


103. Messier: A High-Resolution Corpus for Cross-Benchmark Agent Evaluation ​

Author: Stefan Krsteski, Charlotte Meyer, Guillaume Allegre, Tony O'Halloran, Alexandre Sallinen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.DB

arXiv:2607.25891v1 Announce Type: new Abstract: Evaluating AI agents in interactive environments is hindered by fragmented tasks, scaffolds, verifiers, and scoring rules. Existing efforts focus on narrow settings, remain limited in scale, or require costly reruns, leaving much of the empirical recor...

📖 Read original article


104. Interactive Reward Agent: GUI Task Evaluation via Environment-State Verification ​

Author: Chenrui Shi, Yuwei Wu, Yang Liu, Ruining Feng, Zirui Shang, Zhi Gao, Lifeng Fan, Che Sun
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25904v2 Announce Type: new Abstract: Graphical user interface task evaluation aims to determine whether a GUI agent has successfully completed a user instruction. Automated GUI task evaluation has received increasing attention because the evaluation results can serve as reward signals for...

📖 Read original article


105. Toward Standardized Cross-Vendor Agent Tool Trust Management in Autonomous Networks ​

Author: Ravi Kant Sharma, Ashutosh Uttam, Ajay Kumar
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.NI

arXiv:2607.25914v1 Announce Type: new Abstract: Autonomous Network Levels 4-5 require AI agents to invoke tools across vendor boundaries without human oversight, yet existing management standards lack a standardized mechanism for cross-vendor trust visibility. When a tool from Vendor B is compromise...

📖 Read original article


106. Penelope: Localized Latent Recurrence for Efficient Structured Reasoning ​

Author: Yutong Chen, Shouqian Shi, Xinran Liu, Haochen Wang, Jiaying Wang, Tianxing Xu, Yuanxi Wang, Zirui Ding
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25915v1 Announce Type: new Abstract: Complex structured reasoning tasks often require additional computation, yet current language models obtain it mainly by increasing parameter scale or by serializing intermediate steps as chain-of-thought (CoT) tokens. The former raises training and de...

📖 Read original article


107. dtControl2+$\varepsilon$: Trading Optimality for Explainability in MDPs via Decision Trees ​

Author: Tereza Kinsk'a, Jan K\v{r}et'insk'y, Tobias Meggendorfer, Sabine Rieder, Maximilian Weininger
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.25925v1 Announce Type: new Abstract: Over the past decade, decision trees have been used to represent controllers (a.k.a. policies) in an explainable way, with dtControl2 as a current state-of-the-art tool. However, for systems that are large or have many corner cases, even such represent...

📖 Read original article


108. A Cost-Effective Multimodal LLM Reasoning Framework for Question Answering over Irregular Clinical Time Series ​

Author: Frank Nie, Ethan B Liu, Yuan Zhu, Wei Fan, Jindong Han
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.25947v1 Announce Type: new Abstract: Question answering (QA) over irregular clinical time series (ICTS) plays a pivotal role in a wide range of healthcare applications. Although recent multimodal time-series large language models (LLMs) have shown considerable promise in general-purpose t...

📖 Read original article


109. Large Language Model for Operations Research Formulation Selection in Multi-Warehouse Inventory Allocation ​

Author: Jintao Xu, Yingzheng Ma, Jiong Dong, Yongzhi Qi, Jianshen Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, math.OC

arXiv:2607.25956v1 Announce Type: new Abstract: Multi-warehouse inventory allocation is typically formulated as a mixed-integer programming (MIP) problem, yet no single formulation consistently matches heterogeneous instance-level regimes induced by demand concentration, inventory imbalance, repleni...

📖 Read original article


110. CHARM: A Multimodal Graph Foundation Model with Hierarchical Context Modeling for Zero-Shot Transfer ​

Author: Ankang Yang, Jitao Zhao, Di Jin, Yuxiao Huang, Dongxiao He
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.26023v1 Announce Type: new Abstract: Graph foundation models (GFMs) have emerged as a promising paradigm for transferring knowledge across graph domains and tasks. Real-world graphs associate nodes with text, images, and other modalities, making multimodal graphs essential for representin...

📖 Read original article


111. Falling Behind Drives Unsafe Development in an Idealised AI Race Experiment ​

Author: Elias Fern'andez Domingos, The Anh Han
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.GT, econ.GN, q-fin.EC

arXiv:2607.26034v1 Announce Type: new Abstract: Technological races create tension between speed and safety: actors may gain by moving faster than competitors, even when risky development is harmful. This is prominent in debates about artificial intelligence (AI), where competitive pressure is often...

📖 Read original article


112. Desktop-Delta Bench: Do Computer-Use Models Understand Desktop GUI Transitions? ​

Author: Abhishek Pillai, Samir Kumar Nayak, Yuan Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CV

arXiv:2607.26041v1 Announce Type: new Abstract: Computer-use agents (CUAs) increasingly act through desktop GUIs to complete long-horizon tasks. Current benchmarks primarily measure end-task success or single-frame grounding. Neither isolates whether a model can reconstruct the causal, task-relevant...

📖 Read original article


113. Untrusted Authors, Trusted Answers: A Calculus of Fidelity-Graded Translations ​

Author: Christoph Kirsch
Published: 7/29/2026, 4:00:00 AM
Categories: cs.PL, cs.AI, cs.LO, cs.SE

arXiv:2607.14137v2 Announce Type: cross Abstract: To answer a question about a program, move the program to where the question is decidable. Every such move is a translation, and every translation is a place to be wrong. We study translation as a graph -- many languages, a few reasoning targets, ind...

📖 Read original article


114. Domain-Prior-Regularized Graph Modeling for Anomaly Detection in Cyber-Physical Systems ​

Author: Youngseok Hwang, Joonsung Kwon, Geonwoo Lee, Hyunwoo Park
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.23197v1 Announce Type: cross Abstract: Anomaly detection on multivariate sensor time series is critical for industrial monitoring of cyber-physical systems (CPS), where even subtle deviations from normal behavior can indicate process disruption. Recent graph-based approaches have made sig...

📖 Read original article


115. Neural Network Learning of One-Bit Protocols for Qubit Measurement Simulation ​

Author: Josep Escrig, Mani Zartab, Giulio Gasbarri, Estel Ferrer, Ramon Mu~noz-Tapia, Gael Sent'is
Published: 7/29/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.NA, math.NA

arXiv:2607.23645v1 Announce Type: cross Abstract: Communication complexity provides a natural framework for quantifying the classical resources required to reproduce quantum statistics. In the qubit prepare-and-measure scenario, two classical bits have been shown to be necessary and sufficient to si...

📖 Read original article


116. DocAnnot -- Accelerating the Creation of Key Information Extraction Datasets with GenAI-Powered Auto-annotation ​

Author: Siddartha Reddy, Harikrishnan P M, Goutham Vignesh, Varun V, Vishal Vaddina
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.CV

arXiv:2607.24745v1 Announce Type: cross Abstract: Key Information Extraction (KIE) is vital for many document applications, but creating training datasets is traditionally a time-consuming manual process. We introduce DocAnnot, a framework that significantly accelerates KIE dataset generation. DocAn...

📖 Read original article


117. VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents ​

Author: Seonok Kim
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL

arXiv:2607.24748v1 Announce Type: cross Abstract: Visually-rich documents such as reports, slides, and manuals often distribute the evidence needed to answer a question across multiple pages, mixing text with layout cues, tables, charts, and figures. This work studies multimodal retrieval-augmented ...

📖 Read original article


118. Game AI Not Fun? A Scoping Review and Meta-Analysis on the Differences in Enjoyment between Human and Computer Opponents ​

Author: Ray Ito
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.CY

arXiv:2607.24749v1 Announce Type: cross Abstract: Although advancements in game character AI aim to enhance player engagement, evidence suggests that perceiving an opponent as artificial can diminish the psychological experience. This paper presents a scoping review and meta-analysis of empirical st...

📖 Read original article


119. CARE-MH: Towards Unified, Reproducible, and Comparable Evaluation of Mental Health LLMs ​

Author: Asher Sprigler, Yixue Zhao, Yi Ding
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI

arXiv:2607.24754v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to provide mental health support, requiring reliable evaluation of safety, empathy, and therapeutic appropriateness. However, existing mental health benchmarks are difficult to reproduce and compare ...

📖 Read original article


120. Patterns of Learner-AI Interaction and Academic Performance in an Object-Oriented Programming Course ​

Author: Marina Lepp
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.CY

arXiv:2607.24755v1 Announce Type: cross Abstract: This full research paper examines how different forms of learner-AI interaction relate to learning outcomes in object-oriented programming (OOP) courses. Generative artificial intelligence (GenAI) tools are increasingly used by students in programmin...

📖 Read original article


121. What Gets Lost When Memory Becomes Media? Evaluating AI-Generated Oral History Visualization ​

Author: Kwangsuk Park, Jaehyun Koo, Jiyeon Lee, Anjung Tan, Hyoungchul Park
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI

arXiv:2607.24756v1 Announce Type: cross Abstract: What gets lost when memory becomes media? Diaspora oral-history interviews require a double transformation; first-person recollection to third-person scene, present interview room to past time and place. When generative AI performs this transformatio...

📖 Read original article


122. From Idea to Classroom in Days: Using "Vibe Coding" to Create a Programming Process Visualizer from IDE Activity Logs ​

Author: Heidi Taveter, Marina Lepp
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.CY

arXiv:2607.24757v1 Announce Type: cross Abstract: This paper reports on the rapid development and classroom deployment of a Thonny log visualizer built using AI-assisted ``vibe coding'' to make students' programming processes easily visible to teachers. We developed a web application that analyzes l...

📖 Read original article


123. Verification Without Distrust: Reframing User-Side Oversight as Routine Epistemic Governance in Everyday Human-Chatbot Interaction ​

Author: Aung Pyae
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI

arXiv:2607.24761v1 Announce Type: cross Abstract: Research on human-AI interaction has long framed verification of system outputs as a trust-contingent behavior that better-calibrated trust should reduce. We test this assumption in everyday human-chatbot interaction through a mixed-methods survey of...

📖 Read original article


124. Measuring and Improving Behavioral Consistency in Large Language Models through Fact-Heuristic-Emotion State Enforcement ​

Author: Gi-Hun Lee, Joong Yull Park
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC

arXiv:2607.24765v1 Announce Type: cross Abstract: Large language models (LLMs) can give different answers to the same decision problem across runs, and reverse a decision when their own prior answer returns as context. We ask whether this instability can be measured and partially reduced without cha...

📖 Read original article


125. The Effect of Text Chunk Size on Retrieval-Augmented Generation Performance ​

Author: German Garrido-Lestache Belinchon, Hugo Garrido-Lestache Belinchon
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL

arXiv:2607.24767v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a powerful process for allowing large language models (LLMs) to retrieve relevant information to use as source material during text generation. A critical yet under-explored component of th...

📖 Read original article


126. Three Sides of Retrieval: Factorial Evidence for Document-Side, Query-Side, and Answer-Side Complementarity in RAG ​

Author: Ng S. T. Chong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.24781v1 Announce Type: cross Abstract: RAG systems rely on chunking, which destroys structural information in documents. Existing heading-based retrieval (Jeong et al., 2025) requires multiple LLM calls per document and returns sub-chunks within matched sections. We introduce ToC-guided p...

📖 Read original article


127. DDSNet: Dual-domain Symmetry-aware Network for PCSEL Property Prediction ​

Author: Cen Chen, Haitao Huang, Jiazhi Mao, Feifan Xu, Zhe Zhuang, Yuxiang Ren
Published: 7/29/2026, 4:00:00 AM
Categories: physics.app-ph, cs.AI

arXiv:2607.24785v1 Announce Type: cross Abstract: Efficient exploration of the photonic crystal (PhC) lattice design space is essential for developing photonic crystal surface-emitting lasers. While coupled-wave theory (CWT) provides an effective physical framework, its computational cost remains pr...

📖 Read original article


128. Unlocking Spatial Grounding in Large Audio-Visual Retrieval models ​

Author: Hugo Malard, Michel Olvera, Sanjeel Parekh, Ga"el Richard, Slim Essid, St'ephane Lathuili`ere
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.MM, cs.SD, eess.AS

arXiv:2607.24786v1 Announce Type: cross Abstract: Weak supervision sets a practical regime for audio-visual sound source localization as dense spatial annotations are costly to obtain at scale. The task, however, remains challenging, as models must locate sound sources from temporally aligned audio-...

📖 Read original article


129. From Naive RAG to Deep Agentic Retrieval: An Evolving Context Engineering Pipeline for Regulatory Compliance ​

Author: Mishca de Costa, Muhammad Saleh Anwar, Dave Mercier, Issam Hammad
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL

arXiv:2607.24791v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) is the dominant paradigm for applying large language models (LLMs) to enterprise document corpora, yet naive implementations encounter hard limits as corpus scale and query complexity grow. This paper traces the e...

📖 Read original article


130. AI-Assisted Knowledge Access for Legacy Enterprise Asset Management in Energy Operations: A Practical Retrieval System ​

Author: Dave Mercier, Mishca de Costa, Muhammad Anwar, Mark Randall, Issam Hammad
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.24792v1 Announce Type: cross Abstract: Energy utilities still run engineering work management, engineering procurement, and inventory processes on long-lived enterprise asset management platforms. Replacing these platforms is often cost prohibitive and operationally disruptive, so practic...

📖 Read original article


131. Selective Impairment of Motor Recovery from Typing Errors in Parkinson's Disease: A Survival Analysis ​

Author: Navin Bondade
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.LG

arXiv:2607.24796v1 Announce Type: cross Abstract: Parkinson's disease (PD) affects multiple, dissociable stages of motor and cognitive control. We ask whether passively-collected keystroke dynamics can distinguish two of these stages: noticing a self-generated error (error monitoring) versus recover...

📖 Read original article


132. Reading Without a Reader: Large Language Models Collapse Reading and Writing into a Single Entangled Code ​

Author: Diego Salda~na Ulloa
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.CL, cs.LG

arXiv:2607.24797v1 Announce Type: cross Abstract: In the literate human brain, reading and writing are two doubly-dissociable systems: a ventral decoding route (impaired in pure alexia) and a fronto-parietal encoding route (impaired in pure agraphia), sharing a partial orthographic core. A decoder-o...

📖 Read original article


133. Multimodal Hybrid Retrieval-Augmented Generation for Scientific Document Understanding using Open-Source SLMs ​

Author: Alexandru-Andrei Sauc\u{a}, Ana-Luiza Rusnac
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.24799v1 Announce Type: cross Abstract: Large Language Models tend to hallucinate when answering domain-specific ques tions from scientific documents without prior fine-tuning. Currently, methods such as Retrieval-Augmented Generation partially solve this problem but face different challen...

📖 Read original article


134. When Thinking Before Retrieval Hurts: TraceBound Diagnostics for Adaptive Knowledge-Graph Retrieval ​

Author: Partha Sarathi Purkayastha (ETH Z"urich)
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL

arXiv:2607.24800v1 Announce Type: cross Abstract: Adaptive retrieval promises to make knowledge-graph question answering more robust by letting a controller search, inspect neighborhoods, revise actions, and stop when evidence is sufficient. We study this premise by introducing TraceBound, a lightwe...

📖 Read original article


Author: Yixin Liu, Kang Yin, Hye-Bin Shin, Seong-Whan Lee
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI

arXiv:2607.24806v1 Announce Type: cross Abstract: Error-related potentials (ErrPs) are widely studied neural signatures associated with error processing in human-machine interaction. In realistic settings, error perception often occurs under heterogeneous multisensory feedback, where variability ind...

📖 Read original article


136. A Path Integral Model of Cognition ​

Author: Haruki Emori, Kazunori Kondo, Atsushi Iriki, Andrei Khrennikov
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, quant-ph

arXiv:2607.24807v1 Announce Type: cross Abstract: We develop the mathematical and physical formulation of cognitive cost optimization that underlies the path-integral model of consciousness. The goal-directed cognitive process is modeled as imaginary-time evolution (ITE) under a projector Hamiltonia...

📖 Read original article


137. EEG Emotion Recognition From AI-Generated Biodigital Architecture Images ​

Author: Hongye Yang, Eva Guttmann-Flury
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.HC

arXiv:2607.24808v1 Announce Type: cross Abstract: Emotional responses to biodigital architecture were examined using electroencephalographic (EEG) data from AI-generated images. A pre-experiment involving 336 participants identified 60 images, selected from an initial pool of 600, that elicited stro...

📖 Read original article


138. Retrieval-Augmented Generation in LLMs for Mental Health: Quantifying the Incremental Contribution of Retrieval Within a Layered Safety Architecture ​

Author: Anand Gupta, Akshat Surolia, Shubham Mishra, Shakil Imtiaz, Chaitali Sinha
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL

arXiv:2607.24817v1 Announce Type: cross Abstract: Digital mental health interventions (DMHIs) offer scalable support, but ensuring they accurately detect users' intent during volatile situations can be challenging. Pure parametric Large Language models (LLMs) do not contain specific safety critical ...

📖 Read original article


139. Dual-Level Atomic and Coordination Geometry Learning for Crystal Property Prediction Using Graph Neural Networks ​

Author: Sanjay Chakraborty
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cond-mat.mtrl-sci, cs.AI, cs.CE

arXiv:2607.24818v1 Announce Type: cross Abstract: Accurate prediction of crystal properties remains a key challenge in computational materials science. While graph neural networks (GNNs) such as CGCNN, MEGNet, ALIGNN, and SchNet have shown strong performance, they primarily represent crystals at the...

📖 Read original article


140. Dynamic Multi-Criteria Bottleneck Severity Index (DMBSI) for Semiconductor Wafer Manufacturing: A Genetically Optimised Framework for Reentrant Production Systems ​

Author: Mohammad Sharifur Rahman, Karl McCreadie, Saugat Bhattacharyya, M M Manjurul Islam, Cormac McAteer, Bryan John Baker, Nuala Parker, Girijesh Prasad
Published: 7/29/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG

arXiv:2607.24819v1 Announce Type: cross Abstract: Wafer fabrication exhibits unique characteristics, including reentrant process flows, variable bottlenecks, and highly variable process conditions. In order to identify the most severe bottleneck at each moment in time for semiconductor wafer fabrica...

📖 Read original article


141. Foundation Models for EEG Are Blind to Long-Range Temporal Correlations: A Spectral-Temporal Dissociation Behind Their Cross-Population Fragility ​

Author: Marzieh Zare
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.ET, cs.LG

arXiv:2607.24834v1 Announce Type: cross Abstract: Objective. Electroencephalography (EEG) foundation models (FMs) are trained to reconstruct or contrastively align short patches, then pooled into a fixed embedding. We tested whether these embeddings retained the long-range temporal correlations (LRT...

📖 Read original article


142. MedJudgeRAG: Option-Wise Evidence Judgment with Dynamic Knowledge Graphs for Medical MCQA ​

Author: Seongwon Seo, Seung Hwan Cho, Young-Min Kim
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.24838v1 Announce Type: cross Abstract: In medical multiple-choice question answering (MCQA), Retrieval-Augmented Generation (RAG) can supplement the domain knowledge of language models (LMs). However, since vanilla RAG indiscriminately utilizes retrieved documents, it can degrade LM perfo...

📖 Read original article


143. REPREC: Representation Driven Parameter-Efficient Recommendation System ​

Author: Harshini Kavuru, Dwipam Katariya, Giri Iyengar, Pranab Mohanty, Kalanand Mishra, Kalanand Mishra
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.24845v1 Announce Type: cross Abstract: Large language models (LLMs) have been applied to sequential recommendation by formulating it as a natural language task. Previous work has improved personalization by incorporating collaborative and sequential signals through input conditioning or L...

📖 Read original article


144. Two Views, One Voice: Evidence-Grounded Conversational Music Recommendation ​

Author: Sungwook Yoo, Sewook Yoo
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL

arXiv:2607.24846v1 Announce Type: cross Abstract: Traditional conversational recommenders entangle retrieval and response generation within a single text interface, so exact entity cues fade as the dialogue's intent evolves, which compromises explanation credibility. We address this within the ACM R...

📖 Read original article


145. Extremal Chowla sets and their linear analogues: A human-AI mathematical investigation using Co-Scientist ​

Author: Mohsen Aliabadi, Keith Driscoll, Elliot Krop, Petar Sirkovic, Everett Sullivan, Elahe Vedadi
Published: 7/29/2026, 4:00:00 AM
Categories: math.NT, cs.AI, math.GR

arXiv:2607.24847v2 Announce Type: cross Abstract: We introduce an extremal invariant associated with Chowla-type order conditions in finite groups. A nonempty subset $S$ of a finite group $G$ is called a Chowla set if every element of $S$ has order greater than $|S|$, and we write $C(G)$ for the max...

📖 Read original article


146. Beyond Predictive Accuracy: A Reliability-Aware Audit of Molecular Representations for Human Olfaction ​

Author: Kai Lun Huang (California State University, Fullerton), Wei Chieh Sun (University of Washington)
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG

arXiv:2607.24848v1 Announce Type: cross Abstract: Pretrained molecular encoders are commonly evaluated through downstream prediction, but predictive accuracy alone does not establish that a learned representation captures reproducible scientific structure, adds information beyond strong conventional...

📖 Read original article


147. DisasterTD: Disaster Toponym Disambiguation Using Multimodal LLMs and Cross-View Geolocalization ​

Author: Wenping Yin, Ziqi Liu, Naixia Mou, Weijia Li, Danfeng Hong, Hao Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.24856v1 Announce Type: cross Abstract: Social media imagery (SMI) provides timely and fine-grained ground perspectives that are valuable for situational awareness and emergency response. Unlike satellite or aerial imagery, SMI can capture disaster impacts and ground-level conditions in a ...

📖 Read original article


148. HVM-GraphRAG: Holistic-View Multimodal Graph Retrieval-Augmented Generation on Complex Document ​

Author: Xin He, Yili Wang, Wenqi Fan, Qing Li, Qinggang Zhang, Yi Chang, Xin Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.24861v1 Announce Type: cross Abstract: Question answering (QA) over complex documents requires models to retrieve and integrate evidence distributed across distant document regions and modalities. Multimodal GraphRAG provides a promising direction by organizing document evidence with grap...

📖 Read original article


149. Tokens are All You Need: Dual-purpose Semantic IDs for Achieving LLM-Level I/O Efficiency in recommendation systems ​

Author: Baolei Li, Yiping Yuan, Yilin Zheng, Likang Yin, Ling Liu, Fabio Soldo, Romer Rosales, Xinyang Yi, Lichan Hong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2607.24865v1 Announce Type: cross Abstract: Large-scale recommendation systems face "Memory Wall" bottlenecks due to massive, dense embedding tables. While generative retrieval uses discrete tokens for IDs, high-dimensional context still relies on inefficient dense formats. Inspired by compute...

📖 Read original article


150. GraphRareBench: An Auditable Graph-Evidence Benchmark for Phenotype-Driven Rare-Disease Diagnosis ​

Author: Guiling Guo, Jia Yang, Jiahao Xu, Shuyuan Zheng, Zhonghai Sun, Qiyuan Li
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG

arXiv:2607.24878v1 Announce Type: cross Abstract: Phenotype-driven diagnostic benchmarks usually report the rank of the reference disease, but they rarely reveal which plausible alternatives are ranked above it or what evidence a tool-using model examines before making its decision. We introduce Gra...

📖 Read original article


151. Human Preference aligned Tabular Similarity ​

Author: Frederik Hoppe, Astrid Franz, Marianne Michaelis, Lars Kleinemeier, Udo G"obel
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24880v1 Announce Type: cross Abstract: Task-agnostic tabular embeddings are increasingly used for similarity search in real-world business systems such as Product Lifecycle Management (PLM). However, leading embedding approaches are optimized primarily for prediction tasks - not for produ...

📖 Read original article


152. Agent Retrieval Bench: Evaluating Repository Context Retrieval for Coding Agents ​

Author: Bowen Qin, Yi Xie
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL

arXiv:2607.24882v1 Announce Type: cross Abstract: Modern coding agents are usually evaluated by whether they eventually produce a correct patch, but patch generation depends on an earlier context-acquisition stage: finding the repository files needed for the task. We introduce Agent Retrieval Bench,...

📖 Read original article


153. Beyond "What to Retrieve": Uncertainty in Retrieval-Augmented Code Generation ​

Author: Chandan Kumar Sah, Li Zhang, Xiaoli Lian
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CL, cs.LG

arXiv:2607.24884v2 Announce Type: cross Abstract: Repository-level code generation relies on heterogeneous evidence whose relevance, compatibility, and completeness are inherently uncertain. Similar-code examples, repository context, and project-specific APIs may provide complementary information, b...

📖 Read original article


154. Eliminating Propagation Delay: Attention-Based Spatial-Temporal Fusion Graph Convolution Network for Traffic Flow Prediction ​

Author: Jinpeng Chen, Ziyu Yu, Tao Wang, Jun Ma, Hongbo Gao, Senzhang Wang, Zufeng Zhang, Kaimin Wei
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24885v1 Announce Type: cross Abstract: Predicting traffic flow is crucial to optimizing transportation systems and improving urban mobility. Many graph convolution-based models have been proposed to extract spatial-temporal features and predict traffic flow. However, most focus on spatial...

📖 Read original article


155. Mechanisms of Width Scaling in Normalized Residual Networks: The Effective Alignment Dimension ​

Author: Jinhao Zhang, Zeyu Liu, Zicheng Yan, Yunquan Zhang, Guangming Tan, Fangming Liu, Daning Cheng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24887v1 Announce Type: cross Abstract: Existing theories of neural-network width characterize asymptotic limits, but provide limited guidance on whether an expansion direction identified from finite training data remains beneficial on unseen data. We study this problem for function-preser...

📖 Read original article


156. GAUGE: Grading Agent-Built Financial Models Without a Golden Answer ​

Author: Jiacheng Lu, Sinuo Wang, Wentao Zhao, Rui Sun, Cheng Hua, Tao Song, Hui Cai, Beidi Luan, Zhengze Wu, Lingjing Teng, Yijia He, Jing Li, Daxin Jiang, Zuo Bai, Haibing Guan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE

arXiv:2607.24889v1 Announce Type: cross Abstract: Financial models combine public disclosures with analyst assumptions to produce forecasts and valuations. While some components can be checked mechanically, forecasts, discount rates, and target prices often admit multiple reasonable answers. Existin...

📖 Read original article


157. LLM as Forecasting Planner: Training-Free Text Conditioning for Time-Series Foundation Models ​

Author: Huu Hiep Nguyen, Dung Nguyen, Minh Hoang Nguyen, Dai Do, Hung Le
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24892v1 Announce Type: cross Abstract: Text-conditioned time-series forecasting predicts a series from both its numerical history and natural-language context, allowing forecasts to account for events and constraints that the past alone cannot reveal. This requires both reliable numerical...

📖 Read original article


158. Early Detection of Distributed Backdoors in Multi-Agent LLM Systems: A Characterization Study ​

Author: Diego Fernandez Arias, Dev Prashant Mistry, Ren Wang, Yibo Hu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.AI

arXiv:2607.24893v1 Announce Type: cross Abstract: Multi-agent LLM systems can be attacked by a payload that no single agent ever holds in full: a poisoned tool hides encrypted fragments in its observations, spreads them across several agents, and an external step reassembles and executes them after ...

📖 Read original article


159. Latent Stability Analysis of Malware Representations Under Feature-Space Perturbations ​

Author: Bamidele Ajayi, Ken McGarry
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.AI

arXiv:2607.24896v1 Announce Type: cross Abstract: Static malware detectors are commonly evaluated using clean-sample metrics such as accuracy, F1, ROC AUC, and PR AUC. However, these metrics provide limited insight into how learned malware representations behave when feature vectors are perturbed, h...

📖 Read original article


160. Harm is not Universal: Community-Specific Toxicity Detection is Urgently Needed ​

Author: Xinnuo Xu, Anja Thieme, Daniela Massiceti, Ioana Tanase, Rita Marques, Melanie Fernandez Pradier, Martin Grayson, Camilla Longden, Cecily Morrison
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.24898v1 Announce Type: cross Abstract: State-of-the-art toxicity detectors for text-to-image generation adopt a one-size-fits-all approach: a single universal model applying fixed safety guidelines to all users. Our empirical evidence shows that these detectors fail to shield marginalized...

📖 Read original article


161. Multiclass Classification without Labels via Posterior Simplex Geometry ​

Author: Rapha"el Bonnet-Guerrini, Johann Ioannou-Nikolaides, Troels Petersen, Vincenzo Piuri
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.GA, cs.AI, stat.ML

arXiv:2607.24943v1 Announce Type: cross Abstract: In many classification problems, reliable instance-level labels are unavailable. However, it is often possible to construct weakly enriched unlabeled samples: datasets selected by different cuts, sources, populations, or experimental conditions that ...

📖 Read original article


162. Stable FP4 Training via Transposition-Invariant Block Quantization ​

Author: Mehdi Rahimifar, Amin Darabi, Mehran Taghian Jazi, Xing Huang, Yao Wang, Zhijun Tu, Yufei Cui, Yunke Peng, Hongliang Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24953v1 Announce Type: cross Abstract: Reducing training precision is a key lever for improving the e ciency of large language model (LLM) training, but pushing beyond FP8 to 4-bit oating point (FP4) remains challenging due to instability during optimization. We identify a fundamental sou...

📖 Read original article


163. Generative Distributionally Robust Optimization ​

Author: Ziwei Zhang, Jonathan Yu-Meng Li, Zhihao Jin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC, stat.ML

arXiv:2607.24983v1 Announce Type: cross Abstract: Generative models are increasingly adopted in distributionally robust optimization (DRO), but existing approaches trade off model compatibility and adversarial structure: methods that accept arbitrary samplers do not restrict worst-case laws to a gen...

📖 Read original article


164. Automatic Knowledge Graph Construction and Query for Earthquake Catalogs ​

Author: Yuxin Zhou, Huai Zhang, S. Mostafa Mousavi
Published: 7/29/2026, 4:00:00 AM
Categories: physics.geo-ph, cs.AI, cs.LG

arXiv:2607.24984v1 Announce Type: cross Abstract: In recent years, the number of events in earthquake catalogs has significantly increased due to the utilization of more effective deep learning based detectors and phase pickers but answering open ended questions such as what characterizes this seque...

📖 Read original article


165. Preliminary Guidelines for Using and Evaluating GenAI Tools to Support Systematic Literature Reviews ​

Author: Barbara Kitchenham, Sebasti'an Pizard, Lech Madeyski, Ronnie de Souza Santos, Martin Shepperd, David Budgen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.AI

arXiv:2607.24991v1 Announce Type: cross Abstract: Context: Generative AI (GenAI) and Large Language Models (LLMs) are increasingly used for academic tasks in software engineering and beyond, including systematic literature reviews (SLRs). However, while capable of summarizing text, there is no guara...

📖 Read original article


166. Calibrated Partial Resets: Preventing Policy Collapse in Continual Reinforcement Learning ​

Author: Luc McCutcheon, Evangelos Chatzaroulas, Saber Fallah
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO

arXiv:2607.24996v1 Announce Type: cross Abstract: Neural networks are hindered by accumulating dormant neurons and loss of expressivity throughout training, particularly in non-stationary data settings, such as continual supervised and reinforcement learning. Recently, neuron resets have been used t...

📖 Read original article


167. CogArena: A Multimethod Evaluation of Cognitive Ability Structure in Large Language Models ​

Author: Dengzhe Hou, Lingyu Jiang, Fangzhou Lin, Kazunori D Yamada
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.24999v1 Announce Type: cross Abstract: LLM cognitive scores are increasingly summarized as per-ability profiles whose dimensions should converge across tasks, respond selectively to matched interventions, and generalize beyond the models used to define them. We introduce CogArena, a proce...

📖 Read original article


168. Authoring Agent Skills: A Software-Engineering Approach ​

Author: Giuseppe Destefanis
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.AI

arXiv:2607.25032v1 Announce Type: cross Abstract: Agent Skills are an emerging way to extend large language model agents with reusable procedural knowledge that the agent loads on demand. Anthropic introduced Agent Skills and published the format as an open specification supported across several age...

📖 Read original article


169. Grounded in Consensus, In Step With Emerging Science: A Consensus-Anchored Multi-Corpus Clinical Chatbot for Long COVID ​

Author: Yining Wu, Philip DiGiacomo, Ying Ding, William Brode
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.25038v1 Announce Type: cross Abstract: Long COVID (LC) poses a challenge for clinical decision support because relevant evidence is distributed across sources with different update cycles, evidentiary roles, and levels of clinical maturity. We present a clinician-facing chatbot that organ...

📖 Read original article


170. Extended Reality as a Mediation Layer for Situated Human Control in Human-Robot Teaming ​

Author: Jens Grubert, John Dudley, Eyal Ofek, Per Ola Kristensson
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.RO

arXiv:2607.25047v1 Announce Type: cross Abstract: Extended Reality (XR) is increasingly used in human-robot interaction to communicate robot intent, planned motion, reachability, and state. We argue that XR should also be understood as a mediation layer for situated human control in human-robot team...

📖 Read original article


171. Lantern: Conflict-Aware Gradient Blending for Physics-Guided Diffusion Models in Calorimeter Simulation ​

Author: Farzana Yasmin Ahmad, Vanamala Venkataswamy, Geoffrey Fox
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25060v1 Announce Type: cross Abstract: Monte Carlo simulation of calorimeter showers is a principal bottleneck for the High-Luminosity LHC, and diffusion models have emerged as fast, high-fidelity surrogates. Their denoising objective is purely statistical, however: a model can minimize i...

📖 Read original article


172. DS@GT ARC at CheckThat! 2026: LLM-Based Trace Ranking and Grouped Reward Modeling for Multilingual Numerical Claim Verification ​

Author: Sagnik Sinha, Shreyas Shrestha
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.25069v1 Announce Type: cross Abstract: Automated verification of numerical claims is a challenging problem, as it requires both language understanding and quantitative reasoning. This paper describes our system for CLEF 2026 CheckThat! Task 2, which focuses on ranking reasoning traces gen...

📖 Read original article


173. Spectral Truncation in Synthetic Control ​

Author: Mojtaba Eslami
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ME, cs.AI, cs.LG, econ.EM, stat.AP

arXiv:2607.25074v1 Announce Type: cross Abstract: Synthetic control (SC) matches a treated unit's pre-treatment trajectory to a weighted combination of donor units. We study Spectral SC, which instead matches the treated unit in coordinates defined by the leading temporal singular vectors of the don...

📖 Read original article


174. Evaluating Communicative Belief Updates in Large Language Models via Implicature Recognition and Cancellation ​

Author: Cesare Spinoso-Di Piano, Verna Dankers, Marius Mosbach, Jackie Chi Kit Cheung
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.25094v2 Announce Type: cross Abstract: Human language is driven by unspoken beliefs and belief updates, making these critical to model for successful communication between large language models (LLMs) and their users. In this paper, we evaluate the ability of LLMs to recognize unspoken be...

📖 Read original article


175. OPERA: Offline Policy-guided Expert Routing and Adaptation for Universal Biomedical Image Analysis ​

Author: Zihan Li, Feiyang Liu, Dandan Shan, Ruibo Wang, Qingqi Hong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV

arXiv:2607.25108v1 Announce Type: cross Abstract: Biomedical image analysis spans diverse modalities and tasks, yet real-world deployment is hindered by severe distribution shifts across scanners, protocols, and patient populations. High-performing models consequently require repeated domain-specifi...

📖 Read original article


176. Analysis of the Shortcut Learning and Clever Hans Effect in CNN based ECG Image Classification ​

Author: Abhay Kumar Pathak, Mrityunjay Chaubey, Manjari Gupta, Deepti Mishra
Published: 7/29/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG

arXiv:2607.25117v1 Announce Type: cross Abstract: Deep learning models for ECG image classification may achieve high accuracy by exploiting non-physiological visual cues instead of ECG waveform morphology. Given the black-box nature of deep learning models, their promise of high predictive performan...

📖 Read original article


177. Learning from 53.6K Real-World Developer Edits of AI-Generated Code ​

Author: Jenny T. Liang, Mihika Bairathi, Wayne Chi, Ameet Talwalkar, Nishant Subramani, Valerie Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.HC, cs.LG

arXiv:2607.25130v1 Announce Type: cross Abstract: Imperfections in AI-generated code require that software developers modify the generated code manually, or by re-prompting an AI programming assistant. Manual code edits provide more realistic and granular information on editing behavior than Git com...

📖 Read original article


178. Agentic AI for Scientific Reasoning in Autonomous Quantum Sensing Experiments ​

Author: Takuya Isogawa, Ryotaro Okabe, Nutdech Phadetsuwannukun, Mingda Li, Paola Cappellaro
Published: 7/29/2026, 4:00:00 AM
Categories: quant-ph, cs.AI

arXiv:2607.25145v1 Announce Type: cross Abstract: We implement an agentic AI workflow built around a large language model (LLM) agent for autonomous experiments with nitrogen-vacancy (NV) centers in diamond. NV centers are a widely used platform for quantum sensing, and the ability to control many m...

📖 Read original article


179. OrganLens: Organ-Specific Representation Learning for CT Foundation Models ​

Author: Zhixuan Ge, Anqi Li, Sadeer Al-Kindi, Hanwen Xu, Wei Qiu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25164v1 Announce Type: cross Abstract: A CT examination captures multiple organs, but many biomedical questions concern abnormalities, prognosis, or longitudinal change in a specific organ. These questions require a separate representation for each organ within the same CT volume. Existin...

📖 Read original article


180. CondPSE: A Polynomial-Filtered Structural Encoder with Conditional Modulation for Graphs ​

Author: Woohyun Lee, Hogun Park
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25169v1 Announce Type: cross Abstract: Message-passing graph neural networks are bounded by the 1-WL test and can miss topological structure that distinguishes non-isomorphic graphs. Positional and structural encodings (PSE) inject such topology-derived signals, and learned PSE encoders s...

📖 Read original article


181. TabRank: Chain-of-Thought Distillation for Table Re-Rankers ​

Author: Adarsh Singh, Kushal Raj Bhandari, Jianxi Gao, Soham Dan, Vivek Gupta
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.IR

arXiv:2607.25182v1 Announce Type: cross Abstract: The ability to retrieve relevant tables for answering questions is a key task for structured information retrieval. Multi-stage retrieval systems rely heavily on rerankers to refine candidate lists produced by efficient first-stage retrievers. As a r...

📖 Read original article


182. RIDGE: An Autonomous Framework for Validation and Method Discovery in LLM-Generated Option Pricing ​

Author: Liexin Cheng, Xue Cheng, Shuaiqiang Liu, Cornelis W. Oosterlee
Published: 7/29/2026, 4:00:00 AM
Categories: q-fin.CP, cs.AI

arXiv:2607.25199v1 Announce Type: cross Abstract: Automated code generation is becoming an important tool in quantitative finance, where large language models can generate option pricing implementations directly from mathematical model specifications. Validating such implementations, however, requir...

📖 Read original article


183. VaLiDRec: Variable-Length LLM-Aligned Semantic IDs for Generative Recommendation ​

Author: Shutong Qiao, Wei Yuan, Tong Chen, Hao Wang, Quoc Viet Hung Nguyen, Hongzhi Yin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.25209v1 Announce Type: cross Abstract: Generative recommendation commonly represents items using fixed-length semantic identifiers (SIDs) constructed through clustering and quantization. However, these artificial codes may overcompress item semantics, remain misaligned with pretrained LLM...

📖 Read original article


184. TopoGR: Revealing and Preserving Latent Structure of Semantic ID in Generative Recommendation ​

Author: Ziyu Zheng, Zhengshun Du, Yaming Yang, Bin Tong, Guan Wang, Meng Yan, Ziyu Guan, Wei Zhao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.25216v1 Announce Type: cross Abstract: Semantic ID-based generative recommendation tokenizes each item into a sequence of discrete semantic IDs and predicts the next item by generating semantic IDs. However, existing methods typically regard SIDs as independent discrete symbols, while oft...

📖 Read original article


185. Laplace-PSN-IRT: Uncertainty Quantification for Neural Item Response Theory Models of LLM Benchmarks ​

Author: Juan Francisco, Mandujano Reyes
Published: 7/29/2026, 4:00:00 AM
Categories: stat.AP, cs.AI, cs.LG

arXiv:2607.25257v1 Announce Type: cross Abstract: Item Response Theory (IRT) has recently been proposed as a framework for evaluating large language model (LLM) benchmarks by separating a model's latent ability from the properties of individual benchmark items. Existing neural IRT approaches, includ...

📖 Read original article


186. Structure-aware Relative Policy Optimization for Ranking ​

Author: Yiteng Tu, Weihang Su, Zitao Su, Yiqun Liu, Min Zhang, Qingyao Ai
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.25268v1 Announce Type: cross Abstract: Ranking is a fundamental component of modern information access systems. Reinforcement learning (RL) provides a flexible framework for directly optimizing coarse-grained feedback and system-level objectives defined over the complete ranking list. How...

📖 Read original article


187. Where Steering Signals Come From: Activation Source Selection in Activation Steering ​

Author: Jiaran Ye, Lingxu Ran, Zijun Yao, Chenpeng Wang, Yong Jiang, Lei Hou, Juanzi Li, Liangming Pan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.25270v1 Announce Type: cross Abstract: Activation steering controls language models by adding vectors or features to hidden states at inference time, but the upstream source of these steering signals is often treated as a secondary detail. We study this source choice as activation source ...

📖 Read original article


188. Bridging Compute- and Data-Optimal Pretraining ​

Author: Tian Qin, Kimia Hamidieh, David Alvarez-Melis
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.PF

arXiv:2607.25271v1 Announce Type: cross Abstract: Classical compute-optimal scaling laws assume an unbounded supply of fresh pretraining data, yet pretraining is increasingly entering a regime in which compute grows faster than the availability of high-quality data. We propose Compute-Data (CD) scal...

📖 Read original article


189. ScaleResfusion: Residual Rectified Flow based on Residual Vector Field ​

Author: Zhenning Shi, Chen Xu, Junhao Zhang, Kefei Zhang, Linjie Liu, Zhedong Zheng, Tao Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25275v1 Announce Type: cross Abstract: Real-world Image Restoration (Real-IR) aims to recover high-quality (HQ) images from complex and unknown degradations. Although recent diffusion-based methods have substantially improved perceptual quality, their current designs leave two key challen...

📖 Read original article


190. CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition ​

Author: Lai Wei, Chengqi Li, Jiapeng Li, Ruina Hu, Yue Wang, Weiran Huang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG

arXiv:2607.25294v1 Announce Type: cross Abstract: Real-world tasks often require models to learn from task-specific context rather than relying only on pre-trained knowledge. While recent work has highlighted this capability as context learning, existing evaluations mainly focus on textual contexts....

📖 Read original article


191. Hybrid Analysis for Secure MCP Tool Use in LLM Agents ​

Author: Ping He, Yuexiang Xie, Yaliang Li, Shouling Ji
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.AI

arXiv:2607.25297v1 Announce Type: cross Abstract: The rapid development of large language model (LLM) agents has enabled their broad adoption across diverse real-world tasks. To standardize interactions between LLM agents and external environments, Model Context Protocol (MCP) tools have emerged as ...

📖 Read original article


192. Retraction-Free Optimization over the Stiefel Manifold for the LoRA Fine-Tuning ​

Author: Yuan Zhang, Jiang Hu, Zhijian Lai, Lin Lin, Zaiwen Wen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25299v1 Announce Type: cross Abstract: Optimization over the Stiefel manifold plays a significant role in various machine learning tasks. Existing methods either use the retraction operators, requiring costly orthonormalization for large-scale matrices, or employ landing methods that rely...

📖 Read original article


193. CAST: Game Solvers as Turn-Level Teachers for LLM Agents ​

Author: Yu Wang, Yi-Kai Zhang, Wentao Shi, Ziang Ye, Yuchun Miao, Yueqing Sun, Qi Gu, Xunliang Cai, Lan-Zhe Guo, Han-Jia Ye, Fuli Feng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.25308v1 Announce Type: cross Abstract: Training large language models (LLMs) to act in long-horizon games is a promising step toward generalist decision-making, yet reinforcement learning with verifiable rewards (RLVR) relies on sparse final rewards that reveal little about which decision...

📖 Read original article


194. Balanced Soft mixture-of-expert model for Glaucoma Detection ​

Author: Sai Venkatesh Chilukoti, Krishna Rauniyar, Min Shi, Xiali Hei
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25324v1 Announce Type: cross Abstract: Glaucoma is a group of eye diseases that damage the optic nerve, often caused by elevated intraocular pressure. It is a leading cause of irreversible vision loss and is typically developed slowly and painlessly, making it difficult to notice until si...

📖 Read original article


195. Physics-Informed Neural Operator for Warm-Starting Background-Decomposed and Preconditioned PSFD: Enabling Scalable 3-D EUV Mask Simulation ​

Author: Doyun Kim, Werner Gillijns
Published: 7/29/2026, 4:00:00 AM
Categories: physics.optics, cs.AI, cs.LG, physics.app-ph

arXiv:2607.25330v1 Announce Type: cross Abstract: We present a physics-informed neural operator (PINO) trained with pseudo-spectral frequency-domain (PSFD) equations for electromagnetic (EM) scattering problems in EUV lithography. The Fourier neural operator is factorized into a two-dimensional late...

📖 Read original article


196. Specula: Scaling formal specifications for autonomous model checking of system code ​

Author: Qian Cheng, Saad Mohammad Rafid Pial, Ruize Tang, Yiming Su, Emilie Ma, Finn Hackett, Ivan Beschastnikh, Yu Huang, Tianyin Xu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.DC, cs.OS

arXiv:2607.25333v1 Announce Type: cross Abstract: Specula is a push-button agentic system that generates high-quality formal specifications for large, complex system code and uses the specifications for highly effective model checking and bug finding. Specula employs large language model (LLM) based...

📖 Read original article


197. Every Time I Hire a Linguist, Inference Costs Go Down: On Linguistic Rules as Effective Prompt Compressors ​

Author: Jianfei Ma, Zhaoxin Feng, Emmanuele Chersoni, Si Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.25335v2 Announce Type: cross Abstract: Prompt compression shortens LLM input to reduce inference cost, yet existing methods score token importance through LM forward passes. It remains questionable whether such nuanced, costly token selection is necessary. Compression requires identifying...

📖 Read original article


198. Explainable AI for Chronic Kidney Disease Prediction Using Simulated Federated Learning ​

Author: Md Zahid Hasan Ontor, Md Al Amin, Anik Dev Nath, Bikash Kumar Paul
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25348v1 Announce Type: cross Abstract: Chronic Kidney Disease (CKD), characterized by the gradual loss of kidney function, remains a significant public health challenge. Early detection is crucial for preventing severe complications and enhancing patient outcomes. In this study, Federated...

📖 Read original article


199. Data Quality Profiling at Scale with Progressive Sampling: A Benchmark for Data-Centric AI Pipelines ​

Author: Laure Berti-Equille
Published: 7/29/2026, 4:00:00 AM
Categories: cs.DB, cs.AI, cs.CL

arXiv:2607.25356v1 Announce Type: cross Abstract: Data quality profiling -- computing missing-value rates, duplicate fractions, outlier densities, and functional-dependency violations -- is foundational for data-centric AI pipelines, yet exhaustive scans over millions of rows are prohibitively slow ...

📖 Read original article


200. Raven: High-Recall Sequence Modeling with Sparse Memory Routing ​

Author: Arshia Afzal, Aviv Bick, Eric P. Xing, Volkan Cevher, Albert Gu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25357v1 Announce Type: cross Abstract: Long-context recall in linear-time sequence models highlights a tradeoff in how they write to memory. State-based linear models, such as state-space models (SSMs) and linear Transformers, write densely, updating the entire state for each newly arrive...

📖 Read original article


201. Rethinking Likelihood distributions: Student's t Likelihood Boosts Bayesian Neural Network Performance ​

Author: Pei-Hsuan Hsia, Lars H. Heyen, Arvid Weyrauch, Markus Goetz, Achim Streit, Sebastian Krumscheid, Charlotte Debus
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25376v1 Announce Type: cross Abstract: In Bayesian neural networks (BNNs), variational inference is a widely adopted framework for modeling uncertainty in a distributional way, with the evidence lower bound (ELBO) serving as the standard objective function. Several distributions contribut...

📖 Read original article


202. MARS: Multi-Agent Re-ranking for Repeat-Order Food Delivery Recommendation ​

Author: Jiahao Tian, Zhenkai Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.25420v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in recommender systems, but it is often unclear how much performance can be obtained from strong pre-trained backbones alone when they are placed inside a structured recommendation pipeline. In this ...

📖 Read original article


203. From Dyad to Triad: Eliciting XAI Requirements in Stroke Rehabilitation ​

Author: Param Rajpura, Yogesh Kumar Meena
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.ET

arXiv:2607.25423v1 Announce Type: cross Abstract: Eliciting explainable AI (XAI) requirements from stroke survivors presents a methodological challenge with direct implications for the design of trustworthy brain-computer interfaces for rehabilitation. How can patients and caregivers articulate pref...

📖 Read original article


204. Emergent Latent-State Computation under Stochastic Volatility ​

Author: Xiaoyu Huang, Lulu Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-fin.ST

arXiv:2607.25459v1 Announce Type: cross Abstract: Mechanistic interpretability has largely focused on language models and deterministic toy tasks. Much less is known about how sequence models internally represent latent stochastic dynamics under noisy, partially observed observations. We study this ...

📖 Read original article


205. Seen, Said, or Forgotten? A Causal Audit of Visual KV Memory Across Dialog Turns ​

Author: Hong Chen, Kang Chen, Yuxuan Fan, Bo Wang, Yubo Gao, Yuanlin Chu, Xuming Hu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25467v1 Announce Type: cross Abstract: Stateful multimodal assistants encode an image once but may answer questions about it many turns later. Attention-guided visual-KV eviction assumes that evidence irrelevant now will remain dispensable, although future questions are unknown. We ask wh...

📖 Read original article


206. Architectural Backdoors in Vision-Language Model Supply Chains via Representation Steering ​

Author: Maria Rosaria Briglia, Igor Maljkovic, Antonio Emanuele Cin`a, Luca Oneto, Iacopo Masi, Fabio Roli
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2607.25479v1 Announce Type: cross Abstract: Vision--Language Models (VLMs) are increasingly deployed through a model supply chain in which pretrained checkpoints, architecture definitions, text encoders, and exported computation graphs are distributed by third parties and reused across downstr...

📖 Read original article


207. Automated Numerical Stability Analysis of Deep Learning Operators ​

Author: Xinye Chen
Published: 7/29/2026, 4:00:00 AM
Categories: math.NA, cs.AI, cs.NA

arXiv:2607.25494v1 Announce Type: cross Abstract: Finite-precision arithmetic unavoidably introduces numerical approximation errors. Numerical computations may use insufficient precision or an improper formulation, which leads to numerical instability. In this paper, we introduce the first unified s...

📖 Read original article


208. Beyond Counts: A Distributional Robustness Margin For Pathology Foundation Models ​

Author: Cl'ement Grisi, Jeroen van der Laak, Geert Litjens
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25497v1 Announce Type: cross Abstract: Pathology foundation models are approaching clinical deployment, yet remain vulnerable to systematic non-biological variation across centres. Differences in tissue preparation, staining and scanning are strongly encoded in their representations, enab...

📖 Read original article


209. At-the-Roofline Sparse Tensor Contractions on Vector Processors for Transformer Inference ​

Author: Bowen Wang, Chi Zhang, Diyou Shen, Renzo Andri, Navaneeth Kunhi Purayil, Luca Benini
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AR, cs.AI

arXiv:2607.25504v1 Announce Type: cross Abstract: Fine-grained weight pruning and activation sparsification have emerged as effective approaches for reducing the compute and memory cost of inference for Transformer models. In the moderate-sparsity regime, Gustavson's dataflow provides a natural exec...

📖 Read original article


210. I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models ​

Author: Yimao Guo, Zuomin Qu, Wei Lu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25522v2 Announce Type: cross Abstract: The rapid advancement of video generation models has led to the increasing misuse of image-to-video (I2V) models. Although substantial progress has been made in detecting AI-generated videos, proactive defenses against I2V models remain underexplored...

📖 Read original article


211. ReLATE: Reliability-Guided Evidence Fusion for Robust UAV--Satellite cross-view Geo-Localization ​

Author: Haochen Jiang, Jialei Pan, Yuzhe Sun, Zhe Dong, Lecheng Ren, Yanfeng Gu, Tianzhu Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25524v1 Announce Type: cross Abstract: Unmanned aerial vehicle (UAV)-satellite cross-view geo-localization matches UAV images against satellite imagery and has achieved impressive accuracy on clean (non-degraded) image benchmarks. In real-world flights, however, UAV observations are frequ...

📖 Read original article


212. Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation ​

Author: Weiming Zhuang, Jiabo Huang, Jingtao Li, Zhizhong Li, Chen Chen, Sina Sajadmanesh, Lingjuan Lyu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25527v1 Announce Type: cross Abstract: Unifying visual understanding and generation in one model holds immense promise, but remains challenging and expensive due to heavy compute and data demands and conflicts between the visual features needed for these two capabilities. To address these...

📖 Read original article


213. Multi-Scale Structural Features for Continual, Comprehensible Visual Recognition in a Developmental Learning Framework ​

Author: Zeki Doruk Erden
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.25531v1 Announce Type: cross Abstract: Contemporary machine learning struggles to learn continually, reuse prior knowledge, and expose a comprehensible internal structure. A recently proposed developmental, gradient-free learning framework addresses these limitations by learning a discret...

📖 Read original article


214. Visual prompt engineering for video models ​

Author: Robert Geirhos, Yuxuan Li, Thadd"aus Wiedemer, Neha Kalibhat, Zi Wang, Mani Malek, Oyvind Tafjord, Kevin Swersky, Been Kim, Priyank Jaini
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25537v1 Announce Type: cross Abstract: In the age of foundation models, a model is only as good as its prompt. For this reason, prompt engineering has become an essential technique for improving language model performance. Since video models are currently becoming foundation models for vi...

📖 Read original article


215. Less is More: Modality-Decoupling for General AIGC Audio-Video Detection ​

Author: Jielun Peng, Yabin Wang, Yaqi Li, Jincheng Liu, Xiaopeng Hong, Athanasios V. Vasilakos
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25543v1 Announce Type: cross Abstract: Generative AI has rapidly expanded audio-visual forgery beyond human-centric deepfakes into general scenes. Existing AIGC detection methods assume audio-visual content correspondence, identifying forgeries by spotting cross-modal inconsistencies. How...

📖 Read original article


216. CORF-GS: Real-Time Wireless Radiance Field Reconstruction via Coupled Optical-RF Gaussian Splatting ​

Author: Jinya Zhang, Jiajia Guo, Chao-Kai Wen, Shi Jin
Published: 7/29/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.CV, cs.IT, math.IT

arXiv:2607.25569v1 Announce Type: cross Abstract: Recent advances in 3D Gaussian Splatting (3DGS)-based wireless radiance field (WRF) reconstruction provide an efficient solution for wireless channel modeling. However, existing WRF reconstruction methods rely on pre-collected observations and offlin...

📖 Read original article


217. The LAIA Dataset: Labelled Attention for Intelligent Automobiles ​

Author: A. Contreras, D. Porres, R. Abad, P. Cano, A. Levy, G. Villalonga, A. M. L'opez, A. Hern'andez-Sabat'e
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.SE

arXiv:2607.25570v2 Announce Type: cross Abstract: The development of autonomous vehicles (AVs) usually relies heavily on data-driven artificial intelligence (AI) models that require large volumes of sensor data with ground-truth annotations. While modular architectures are widely used, end-to-end dr...

📖 Read original article


218. IRIS: Reusable Identity Representations from Frozen LLMs for Entity Alignment ​

Author: Xinran Liu, Shengtao Li, Shouqian Shi, Ge Wang, Xin-Wei Yao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.25579v1 Announce Type: cross Abstract: Entity alignment (EA) identifies entities across knowledge graphs (KGs) that refer to the same real-world object. Conventional EA methods mainly exploit explicit graph structures and textual fields, which often provide insufficient semantic understan...

📖 Read original article


219. Beyond Self-Knowledge: Propagating Uncertainty Across Reasoning and Retrieval in LLMs ​

Author: Chandan Kumar Sah, Li Zhang, Xiaoli Lian
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL

arXiv:2607.25600v2 Announce Type: cross Abstract: Retrieval-augmented generation improves knowledge-intensive question answering, but indiscriminate retrieval can introduce irrelevant evidence and unnecessary computation. We investigate whether verbalized confidence from black-box language models ca...

📖 Read original article


220. Physics-Informed Broad Learning System: An Efficient Backpropagation-Free Framework for Solving Partial Differential Equations ​

Author: Pinki Khatun, M. Sajid, Abhinav Jha, M. Tanveer
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25608v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful paradigm for solving partial differential equations (PDEs) by embedding governing physical laws into deep neural networks. However, their reliance on computationally expensive gradie...

📖 Read original article


221. Contrastive Representation Learning of Longitudinal Disease Trajectories on Temporal Graphs ​

Author: Bastian Pfeifer
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.QM

arXiv:2607.25609v1 Announce Type: cross Abstract: Understanding disease trajectories from longitudinal clinical data remains challenging due to complex temporal dynamics and heterogeneous patient cohorts. Here, we present a contrastive representation learning framework that models multivariate disea...

📖 Read original article


222. A Human-in-the-Loop Corpus for LLM-Based Simplification of Scientific Summaries ​

Author: Kyuri Im, Michael F"arber
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC

arXiv:2607.25630v1 Announce Type: cross Abstract: Interdisciplinary research is accelerating, yet scientific papers remain difficult to understand outside their home fields. We study large language model (LLM)-based simplification of scientific texts and present a human-in-the-loop workflow that tra...

📖 Read original article


223. Construction-Driven Injection: Linguistically-Grounded Edit-Based Code-Mixing Fingerprints for Large Language Models ​

Author: Yongyi Cui, Yue Li, Tianbao Jiang, Xin Yi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.25633v1 Announce Type: cross Abstract: Large language models (LLMs) are costly intellectual assets that remain exposed to unauthorized redistribution and commercial misuse. Injected fingerprints, i.e., trigger--target pairs embedded in model behavior, offer a practical, black-box-verifiab...

📖 Read original article


224. F(AI)2R: Who Did What, and Who Checked? Verifiable AI Provenance as an Executable Skill ​

Author: Florian Krebs
Published: 7/29/2026, 4:00:00 AM
Categories: cs.DL, cs.AI, cs.SI

arXiv:2607.25637v1 Announce Type: cross Abstract: F(AI)2R is FAIR research with AI in the loop, twice: an AI-assisted authoring pass and a machine-readable audit pass over every artefact. AI systems now draft, refactor, and verify research artefacts, yet their contributions are rarely recorded in a ...

📖 Read original article


225. OmniPhys: Knowledge-Graph-Driven Benchmarking and Collective Optimization for Physical Commonsense in Text-to-Image Generation ​

Author: Yajing Xu, Yarong Lan, Jiaoyan Chen, Yichi Zhang, Jeff Z. Pan, Mingchen Tu, Zhizhen Liu, Wen Zhang, Huajun Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25641v1 Announce Type: cross Abstract: While text-to-image models exhibit remarkable visual fidelity, they frequently violate fundamental physical commonsense. Existing benchmarks often rely on coarse-grained descriptions, failing to diagnose the mastery of specific physical principles. M...

📖 Read original article


226. KQFuzz: Knowledge-Guided Fuzzing for Quantum Libraries via Large Language Models ​

Author: Fuyuan Xia, Qixin Zhang, Chenhao Ying, Haojin Zhu, Shuai Wang, Yuan Luo, Pingchuan Ma, Yuxuan Du
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.MA, quant-ph

arXiv:2607.25647v1 Announce Type: cross Abstract: As quantum computing continually improves, ensuring the reliability and correctness of quantum libraries has become increasingly critical. To this end, many LLM-based fuzzing approaches towards quantum libraries have been proposed to uncover potentia...

📖 Read original article


227. Why Public Service AI Governance Frameworks Risk Failing in the Age of General-Purpose AI: Lessons from Policing ​

Author: Sam Relins, Daniel Birks
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CY, cs.AI

arXiv:2607.25648v1 Announce Type: cross Abstract: Public services face growing pressure to adopt artificial intelligence (AI) to close the gap between rising demand and falling resources. That pressure has intensified with general-purpose AI (GPAI): AI built on large language models that can be dire...

📖 Read original article


228. MyMentorLLM: A psychotherapy GenAI environment with multimodal voice/text patients, trainees and experts for deliberate practice ​

Author: Rodolfo Rizzi, Alessandro Grecucci, Massimo Stella
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.25667v1 Announce Type: cross Abstract: Psychotherapists need repeated training and supervision by experts; however, scalability is problematic. Here we present MyMentorLLM, a multimodal voice- and text-based simulation environment for deliberate practice, used to generate 2,100 complete C...

📖 Read original article


229. DynaBridge: Dynamic Summary-Guided Cross-Task Multimodal Fusion for DASS-Structured Mental Health Assessment ​

Author: Shiyu Teng, Haichen Yu, Jiaqing Liu, Hao Sun, Yu Song, Shurong Chai, Ruibo Hou, Lanfen Lin, Yen-Wei Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MM

arXiv:2607.25679v1 Announce Type: cross Abstract: Multimodal behavioral analysis offers a scalable approach to assessing depression, anxiety, and stress, yet generic fusion models often ignore the psychometric structure of questionnaire labels. In DASS-21, risk labels are derived from ordered sympto...

📖 Read original article


230. Rashomon Alignment ​

Author: Mois'es Santos, Peter van der Putten, Bernhard Pfahringer, Carlos Soares
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25680v1 Announce Type: cross Abstract: We propose Rashomon Alignment (RA), a new measure to assess functional similarity between two models. Existing functional similarity measures are distributional, quantifying differences between outputs of models applied to real-world data. However, t...

📖 Read original article


231. From Deterministic to Generative Deep Learning for Urban Air Quality Reconstruction from Sparse Observations ​

Author: Abhishek A. Sabnis, Mihai Mitrea, Lya Lugon, Karine Sartelet, Marc Bocquet, Xiaoyuan Cheng, Shupeng Zhu, Sibo Cheng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25687v1 Announce Type: cross Abstract: Full-field reconstruction of air pollution is essential for evaluating pollution exposure and supporting public health decision-making. However, the complex interactions among pollutants, hard-to-predict weather patterns, and limited monitoring stati...

📖 Read original article


232. Tools Are Not Islands: Set-Level Tool Retrieval for LLM Agents via Query-Conditioned Hyperedge Prediction ​

Author: Xinyi Hong, Pinjun Dong, Xinyang Yu, Binyan Jiang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR

arXiv:2607.25718v2 Announce Type: cross Abstract: Large language model (LLM) agents increasingly rely on invoking external tools to complete real-world tasks. Tool retrieval, which selects a small task-relevant subset from a library of thousands of tools before the agent acts, has therefore become a...

📖 Read original article


233. Shared Voxel-Map-Based Cooperative Indoor UAV Guidance with a Multi-Agent Soft Actor-Critic Controller ​

Author: Thomas Hickling, Dylan Wynne, Yu Su, Nabil Aouf
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.25728v1 Announce Type: cross Abstract: This paper presents a cooperative indoor UAV guidance framework that combines a shared voxel-map world model with a multi-agent Soft Actor-Critic (MASAC) controller. Multiple drones fuse 360 LiDAR observations into a common world-frame occupancy map,...

📖 Read original article


234. Image Quality Dependent Degradation for AI Systems ​

Author: Yannick Kees, Elena Hoemann, Frank K"oster, Sven Hallerbach
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25736v1 Announce Type: cross Abstract: Perception is one of the primary applications where neural networks outperform conventional algorithms. One example is AI systems for automated driving, which can detect pedestrians based on image data and avoid them accordingly. A substantial challe...

📖 Read original article


235. SpectONet: A Physics-Guided Spectral Deep Operator Network for Euler-Bernoulli Beam Dynamics ​

Author: Shivani Saini, Ramesh Kumar Vats, Arup Kumar Sahoo
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DS

arXiv:2607.25790v1 Announce Type: cross Abstract: This paper proposes a novel physics-guided spectral deep operator network, termed SpectONet, for solving Euler-Bernoulli beam (EBB) vibration problems. The proposed framework integrates the operator-learning capability of DeepONet with physics-inform...

📖 Read original article


236. Lowering the implementation barrier of neutral-atom quantum computing with agentic workflows ​

Author: Constantin Dalyac, Alexandre Dauphin, Lo"ic Henriet, Christophe Jurczak
Published: 7/29/2026, 4:00:00 AM
Categories: quant-ph, cond-mat.quant-gas, cs.AI

arXiv:2607.25834v1 Announce Type: cross Abstract: Quantum computers are moving from research laboratories to industrial machines accessible via the cloud and integrated into high-performance computing facilities. However, translating theoretical quantum protocols into hardware experiments remains a ...

📖 Read original article


237. OmniQEC: discovering practical quantum error-correcting codes by an AI scientist ​

Author: Ge Yan, Shanchuan Li, Pengyue Ma, Qixin Zhang, Pingchuan Ma, Jianping Wang, Min-Hsiu Hsieh, Yuxuan Du
Published: 7/29/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.MA

arXiv:2607.25865v1 Announce Type: cross Abstract: Quantum error correction (QEC) is indispensable for scalable fault-tolerant quantum computing. However, discovering QEC codes that remain effective is challenging, as logical performance depends on the interplay between code structure, hardware, synd...

📖 Read original article


238. How Do LLMs Read Bug Reports? An Empirical Study of Attention in LLMs for Automated Program Repair ​

Author: Ramtin Ehsani, Irene Manotas, Saurabh Pujar, Luca Buratti, Preetha Chatterjee
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.AI

arXiv:2607.25873v1 Announce Type: cross Abstract: Large Language Model (LLM)-based Automated Program Repair systems are advancing rapidly, yet their performance remains inconsistent. Even when provided with the same contextual information, an LLM may generate a correct patch for one bug but fail on ...

📖 Read original article


239. A2TTA: Anchored-and-Agile Test-Time Adaptation for Evolving Traffic Sensor Networks ​

Author: Du Yin, Xiachong Lin, Yue Tan, Jinliang Deng, Estrid He, Hao Xue, Flora D. Salim
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25875v1 Announce Type: cross Abstract: Traffic forecasting is important for efficient traffic management and route planning in smart cities. Existing traffic forecasting studies typically assume fixed sensor graphs, overlooking the continuous evolution of real-world traffic networks, e.g....

📖 Read original article


240. Stemma: Induced Decision Regions Reveal LLM Provenance ​

Author: Keyu Zhang, Vadim Safronov, Andrew Martin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL

arXiv:2607.25880v1 Announce Type: cross Abstract: LLM provenance testing asks whether a suspect LLM belongs to the same lineage as a source. Existing black-box methods largely infer this relationship from response-level characteristics, but these characteristics may shift under adaptation or deploym...

📖 Read original article


241. A Machine-Learning-Based Gas Lift Optimization Workflow for Unconventional Fields ​

Author: Sha (Sasha), Miao, Alexandra Vendetti, Logan Smart, Gunta Chomchalerm, Yang Chen, Christopher Frazier, Dustin Haralson, Jeremy Sorenson, Xiao Ma, Huafei Sun, Aaron Shinn, Haining Zheng, Xiao-Hui Wu, Peng Xu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE

arXiv:2607.25885v1 Announce Type: cross Abstract: In this paper, we present an automated data-driven workflow using Machine Learning (ML) for gas lift optimization in unconventional fields. This workflow integrates a ML model that accurately forecasts the Gas Lift Performance Curve, and a Bayesian O...

📖 Read original article


242. Device Invariance using Domain Adaptation on Acoustic Scene Classification ​

Author: Abhishek dileep, Shubham Sharma, Padmanabhan Rajan
Published: 7/29/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.SD

arXiv:2607.25887v1 Announce Type: cross Abstract: This paper explores the effectiveness of domain adaptation techniques when using convolutional neural network (CNN)-based and transformer-based feature representations for acoustic scene classification. Two well-known domain adaptation techniques, na...

📖 Read original article


243. Depression Markers in Speech: An Approach based on Tract Variables Dynamics ​

Author: Sahar Altalhi, Tanaya Guha, Alessandro Vinciarelli
Published: 7/29/2026, 4:00:00 AM
Categories: eess.AS, cs.AI

arXiv:2607.25888v1 Announce Type: cross Abstract: This study identifies new depression biomarkers based on the dynamical properties of tract variables, which represent geometric features describing the configuration of the speech articulators. A key advantage of this approach lies in its ability to ...

📖 Read original article


244. Minimizing Targeted Activations: Input-Only Suppression of Evaluation-Awareness Latents in Large Language Models ​

Author: Deepanshu Mody, Samarth Agarwal, Utkarsh Mittal, Dipesh Mahato
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.25907v1 Announce Type: cross Abstract: Activation steering controls model behavior by editing internal activations at inference time. We study its input-side dual: optimizing a fluent prompt so that a chosen internal latent is driven toward zero, with no inference-time model access. Our t...

📖 Read original article


245. AnnoBench: A Benchmark for Visualization Annotation Generation ​

Author: Md Rahat-uz-Zaman, Md Dilshadur Rahman, Andrew McNutt, Paul Rosen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI

arXiv:2607.25911v1 Announce Type: cross Abstract: Annotation is among the most demanding visualization tasks to automate, as it simultaneously requires correctly navigating visual, semantic, and stylistic constraints. Failure to meet any of these conditions severely undermines the utility of an anno...

📖 Read original article


246. SAM3D-Guided Object-Centric Representation Alignment for Vision-Language-Action Models ​

Author: Zonghe Liu (University of Hong Kong), Shanyuan Jie (Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences), Xiaoquan Sun (Huazhong University of Science and Technology), Chen Cao (University of Hong Kong), Zetian Xu (University of Hong Kong), Zongsheng Liu (Beijing University of Aeronautics and Astronautics), Jiayu Chen (University of Hong Kong, Infiforce)
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI

arXiv:2607.25912v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general robot manipulation, but most existing models rely on 2D visual-language backbones and lack fine-grained 3D understanding of target objects, especially under occlusion, pose v...

📖 Read original article


247. Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA ​

Author: Carlos Celemin, Benedict Wilkins, Adri'an Barahona-R'ios, Saman Zadtootaghaj, Nabajeet Barman
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25921v1 Announce Type: cross Abstract: In this work, we study the use of Vision-Language Models (VLMs) for anomaly detection in an agent-driven game Quality Assurance (QA) pipeline focusing on geometry clipping. In this evaluation, a custom exploration agent navigates a game level to coll...

📖 Read original article


248. Face De-Identification: A Domain-Centric Survey from Capture to Processing ​

Author: Hui Wei, Hao Yu, Guoying Zhao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25926v1 Announce Type: cross Abstract: Face de-identification (De-ID) aims to remove or conceal personally identifiable facial features in images or videos to prevent identity recognition while preserving utility for downstream tasks. With the rising emphasis on data privacy and responsib...

📖 Read original article


249. Evaluating Multi-Turn Multimodal Diagnostic Reasoning on Challenging Real-World Clinical Cases ​

Author: Rui Yang, Weihao Xuan, Yi Lin, Zhuhan Bao, Jonathan Chong Kai Liew, Matthew Yu Heng Wong, Nicol'as Lescano, Nikita R. Paripati, Emily Ling-Lin Pai, Jiarui Liu, Heli Qi, Heng-Jui Chang, Benny Kai Guo Loo, Huitao Li, Kunyu Yu, Yufan Wang, Chuan Hong, Shijian Lu, Douglas Teodoro, Naoto Yokoya, Ross Koppel, Mona Diab, Hua Xu, David W. Bates, Nan Liu, Yifan Peng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.25933v1 Announce Type: cross Abstract: Clinical diagnostic evaluation should not only assess whether models can provide correct diagnoses, but also reflect the realities of clinical practice, including progressive disclosure of multimodal information, dynamic updating of diagnostic hypoth...

📖 Read original article


250. MODUS: Decoder-Only Any-to-Any Modeling of Diverse Modalities ​

Author: Mingqiao Ye, Zhaochong An, Zhitong Gao, Xian Liu, Fran\c{c}ois Fleuret, Chuan Li, Amir Zadeh, Serge Belongie, Afshin Dehghan, Jesse Allardice, David Mizrahi, O\u{g}uzhan Fatih Kar, Roman Bachmann, Amir Zamir
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.25948v1 Announce Type: cross Abstract: Any-to-any models predict any modality from any combination of others within a single network, a formulation used in multimodal vision and vision-language models, and increasingly in scientific domains such as ecology and astronomy. Existing any-to-a...

📖 Read original article


251. Detecting Knowledge Inconsistencies Across Text, Tables, and Knowledge Graphs ​

Author: Fanfu Wei, Thibault Ehrhart, Rapha"el Troncy
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.25959v2 Announce Type: cross Abstract: Wikipedia and Wikidata are widely used for information access, LLM pre-training, and retrieval-augmented generation. Their knowledge is deeply connected but scattered across text, tables, and knowledge graphs. This raises a practical question: when t...

📖 Read original article


252. Knowledge-Guided Multimodal Reasoning over Interacting Streams for Video-Level Ambivalence and Hesitancy Recognition ​

Author: Podakanti Satyajith Chary, Barath Parthiban, Pranesh Velmurugan, Adeeba Khan, Nagarajan Ganapathy
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.25961v2 Announce Type: cross Abstract: Ambivalence and hesitancy (A/H) are conflicting affective states that precede the delay or abandonment of health behaviour change. Recognition of A/H at the video level is difficult, since the signal arises from disagreement across and within facial,...

📖 Read original article


253. Reinforcement Learning for Code Optimization ​

Author: Pierre Chambon, Kunhao Zheng, Juliette Decugis, Benoit Sagot, Gabriel Synnaeve
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25970v1 Announce Type: cross Abstract: RL for code correctness is now established: have the model generate a program, run it against hidden test cases, and reward solutions that pass. Extending this to code optimization seems straightforward: just add execution time to the reward. But in ...

📖 Read original article


254. MemLens: A Value-Aware Memory Management System with Interactive Analytics for LLM-based Agents ​

Author: Shuyue Wei, Chang Liu, Zimu Zhou, Yongxin Tong, Lizhen Cui
Published: 7/29/2026, 4:00:00 AM
Categories: cs.DB, cs.AI

arXiv:2607.25992v1 Announce Type: cross Abstract: Recently, memory management has become a key infrastructure for LLM-based agents, as it directly affects long-horizon reasoning, personalized responses, and knowledge reuse. However, existing LLM memory systems typically adopt a coarse-grained (utili...

📖 Read original article


255. Does Runtime Topology Context Improve LLM-Generated Kubernetes Security Patches? ​

Author: Farooq Shaikh
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.AI

arXiv:2607.25995v1 Announce Type: cross Abstract: Kubernetes is central to the cloud-native ecosystem, orchestrating containerised workloads. Recent work suggests that large language models (LLMs) can automate cluster security remediation, generating configuration patches from Kubernetes Security Po...

📖 Read original article


256. Empirical Evaluation of Out-Of-Distribution Performance of Tabular Foundation Models ​

Author: Malena Loza, David Chushig-Muzo, Eva Milara, Luis Bote-Curiel, Luis Estrada-Petrocelli, Felipe Grijalva
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.26000v1 Announce Type: cross Abstract: Tabular Foundation Models (TFMs) have emerged as novel approaches for tabular predictive tasks, demonstrating competitive predictive performance to ensemble tree-based models. Most TFMs are trained and evaluated on independent and identically distrib...

📖 Read original article


257. Pictura: Perspective-View Self-Play at Scale for Driving ​

Author: Yuan Yin, Elias Ramzi, Marc Lafon, Valentin Charraut, Victor Bares, Yihong Xu, 'Eloi Zablocki, Alexandre Boulch, Thibault Buhet, Andrei Bursuc, Matthieu Cord
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.RO

arXiv:2607.26005v1 Announce Type: cross Abstract: Self-play in simulation produces robust driving policies at scale. Demonstrations of such behavior have been made using privileged vectorized observations such as exact poses and velocities, even for occluded agents. This assumes that perception is s...

📖 Read original article


258. MDTransformer: A Hardware-Software Co-Design of Mode-Division Photonic Transformer Accelerator with Inverse-Designed Coherent Crossbar ​

Author: Solomon Micheal Serunjogi, Rachmad Vidya Wicaksana Putra, Ayat Taha, Muhammad Shafique, Mahmoud Rasras
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AR, cs.AI, cs.DC

arXiv:2607.26016v1 Announce Type: cross Abstract: Recently, photonic transformer accelerators (PTAs) have successfully achieved significant speedup and energy efficiency improvements over electronic accelerators for expediting Transformer inference. However, state-of-the-art rely on expensive multi-...

📖 Read original article


259. $\pi\mathbf{R}^2$: Reactive Real-time Flow Policies ​

Author: Sungjae Park, Shubham Tulsiani
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.26055v1 Announce Type: cross Abstract: Generalist manipulation policies increasingly take the form of action-chunking flow policies built on large pretrained backbones. Such chunks run open-loop, so the policy cannot react to sensory input arriving mid-execution, sacrificing \emph{reactiv...

📖 Read original article


260. Pass the Baton: Trajectory-Relayed On-Policy Distillation ​

Author: Haolei Xu, Xiaowen Xu, Haiwen Hong, Zixuan Ni, Hongxing Li, Yiwen Qiu, Weiming Lu, Yongliang Shen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.26057v1 Announce Type: cross Abstract: On-policy distillation (OPD) grounds token-level supervision in the student's own trajectory, yet suffers from prefix failure: once the student commits to a wrong reasoning direction, all subsequent generation builds on this deviation, producing misd...

📖 Read original article


261. Diffusion Model-based Parameter Estimation in Dynamic Power Systems ​

Author: Feiqin Zhu, Dmitrii Torbunov, Zhongjing Jiang, Tianqiao Zhao, Amirthagunaraj Yogarathnam, Yihui Ren, Meng Yue
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.SY, eess.SY

arXiv:2411.10431v3 Announce Type: replace Abstract: Parameter estimation, which represents a classical inverse problem, is often ill-posed as different parameter combinations can yield identical outputs. This non-uniqueness presents a critical barrier to accurate and unique identification. Here we i...

📖 Read original article


262. Real-time Spatial Retrieval Augmented Generation for Urban Environments ​

Author: David Nazareno Campo, Javier Conde, 'Alvaro Alonso, Gabriel Huecas, Joaqu'in Salvach'ua, Pedro Reviriego
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2505.02271v2 Announce Type: replace Abstract: The proliferation of Generative Artificial Ingelligence (AI), especially Large Language Models, presents transformative opportunities for urban applications through Urban Foundation Models. However, base models face limitations, as they only contai...

📖 Read original article


263. Towards Embodied Cognition in Robots via Spatially Grounded Synthetic Worlds ​

Author: Joel Currie, Gioele Migno, Enrico Piacenti, Maria Elena Giannaccini, Patric Bach, Davide De Tommaso, Agnieszka Wykowska
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.RO

arXiv:2505.14366v2 Announce Type: replace Abstract: We present a conceptual framework for training Vision-Language Models (VLMs) to perform Visual Perspective Taking (VPT), a core capability for embodied cognition essential for Human-Robot Interaction (HRI). As a first step toward this goal, we intr...

📖 Read original article


264. On the Design and Evaluation of Human-centered Explainable AI Systems: A Systematic Review and Taxonomy ​

Author: Aline Mangold, Juliane Zietz, Susanne Weinhold, Sebastian Pannasch
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2510.12201v2 Announce Type: replace Abstract: As AI becomes more common in everyday living, there is an increasing demand for intelligent systems that are both performant and understandable. Explainable AI (XAI) systems aim to provide comprehensible explanations of decisions and predictions. A...

📖 Read original article


265. Controllable LLM Reasoning via Sparse Autoencoder-Based Steering ​

Author: Yi Fang, Wenjie Wang, Mingfeng Xue, Boyi Deng, Fengli Xu, Dayiheng Liu, Fuli Feng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2601.03595v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) exhibit human-like cognitive reasoning strategies (\eg backtracking, cross-verification) during the reasoning process, which improves their performance on complex tasks. Currently, reasoning strategies are autonomously...

📖 Read original article


266. JobMatchAI-An Intelligent Job Matching Platform Using Knowledge Graphs, Semantic Search and Explainable AI ​

Author: Mayank Vyas, Abhijit Chakraborty, Vivek Gupta
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2603.14558v3 Announce Type: replace Abstract: Recruiters and job seekers rely on search systems to navigate labor markets, making candidate matching engines critical for hiring outcomes. Most systems act as keyword filters, failing to handle skill synonyms and nonlinear careers, resulting in m...

📖 Read original article


267. DSevolve: Enabling Real-Time Adaptive Scheduling on Dynamic Flexible Job Shop with LLM-Evolved Heuristic Portfolios ​

Author: XinLei Zhou, Jin Huang, Jie Yang, Xinyu Li, Liang Gao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2603.27628v2 Announce Type: replace Abstract: In dynamic flexible job shops, order arrivals, machine breakdowns, and processing-time deviations continually reshape the scheduling state and the priority trade-offs behind dispatching decisions. Dispatching rules are well suited to this setting b...

📖 Read original article


268. The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem ​

Author: Till Mossakowski, Helena Esther Grass
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2604.14990v2 Announce Type: replace Abstract: The prospect of Artificial General Intelligence (AGI) is increasingly driving institutional decisions, and alignment of AGI is a hard problem. The currently dominant AI alignment strategies like reinforcement learning with human feedback or constit...

📖 Read original article


269. Why Does Grounding Hurt Medical VQA? Benchmarking, Diagnosis, and Fine-Tuning of Vision-Language Models ​

Author: Xupeng Chen, Binbin Shi, Chenqian Le, Qifu Yin, Lang Lin, Haowei Ni, Ran Gong, Panfeng Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2604.27720v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly applied to medical visual question answering (Med-VQA), yet whether they can \emph{localize} the evidence behind their answers---a prerequisite for clinical auditability---is poorly characterized. We s...

📖 Read original article


270. The Scaling Properties of Implicit Deductive Reasoning in Transformers ​

Author: Enrico Vompa, Tanel Tammet
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CC, cs.LO, cs.SC

arXiv:2605.04330v2 Announce Type: replace Abstract: We investigate the scaling properties of implicit deductive reasoning over Horn clauses in depth-bounded Transformers. By systematically decorrelating provability from spurious features and enforcing algorithmic alignment, we find that in sufficien...

📖 Read original article


271. AlphaCrafter: Harnessing Multi-Agent Workflows for Cross-Sectional Quantitative Trading ​

Author: Yishuo Yuan, Jiayi Sheng, Sirui Zeng, Jiaqi Wang, Jiaheng Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2605.05580v2 Announce Type: replace Abstract: Quantitative trading agents have demonstrated substantial promise in automating factor discovery, signal aggregation, and portfolio execution. However, existing agent-based trading systems predominantly rely on loosely specified natural-language wo...

📖 Read original article


272. Sheet As Token: A Graph-Enhanced Representation for Multi-Sheet Spreadsheet Understanding ​

Author: Yiming Lei, Yuhang Yao, Yujia Zhang, Yiqi Wang, Bo Guan, Depei Zhu, Chunhui Wang, Zhuonan Hao, Tianyu Shi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2605.05811v2 Announce Type: replace Abstract: Workbook-scale spreadsheet understanding is increasingly important for language-model-based data analysis agents, but remains challenging because relevant information is often distributed across multiple sheets with heterogeneous schemas, layouts, ...

📖 Read original article


273. From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World ​

Author: Pedro Conde, Henrique Branquinho, Valerio Mazzone, Bruno Mendes, Andr'e Baptista, Nuno Moniz
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CR

arXiv:2605.10834v3 Announce Type: replace Abstract: AI pentesting agents are increasingly credible as offensive security systems, but current benchmarks still provide limited guidance on which will perform best in real-world targets. Existing evaluation protocols assess and optimize for predefined g...

📖 Read original article


274. RoboPIN: Grounded Embodied Reasoning via Pinned Chain-of-Thought ​

Author: Yaoting Huang, Yifu Yuan, Linqi Han, Chengwen Li, Shuoheng Zhang, Xianze Yao, Hongyao Tang, Yan Zheng, Jianye Hao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2606.15753v4 Announce Type: replace Abstract: Embodied reasoning requires models to perceive task-relevant objects and spaces in physical environments and maintain consistent visual grounding throughout multi-step reasoning. However, current vision-language models rely on text-only or coordina...

📖 Read original article


275. Psychological Competence as a Missing Dimension in AI Evaluation ​

Author: Marcos Economides, Paul M. Sacher, Samuel Salzer, Alexis Michelle Abellar, Fendi Tsim, Antoine Ferr`ere
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08285v2 Announce Type: replace Abstract: Current AI evaluation frameworks focus primarily on technical performance, including accuracy, robustness, reasoning ability, and policy compliance. These measures remain essential, but they are not sufficient for systems that interact directly wit...

📖 Read original article


276. Neuro-Symbolic Meta-Policies for Temporal Knowledge-Graph Memory under Partial Observability ​

Author: Taewoon Kim, Vincent Fran\c{c}ois-Lavet, Michael Cochez
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.18368v3 Announce Type: replace Abstract: Partially observable reinforcement learning requires deciding what to retain, retrieve, and forget over time. We introduce a neuro-symbolic meta-policy that learns which symbolic memory heuristic to apply at each decision point while keeping execut...

📖 Read original article


277. EviDAG: Auditable Causal DAG Authoring with Biomedical Literature ​

Author: Yi-han Sheu, Michael R. Steigman, Yu Zhou, Bo Wang, Fan-Yu Yen, Jordan W. Smoller
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.21859v2 Announce Type: replace Abstract: Constructing causal directed acyclic graphs (DAGs) is a core step in biomedical causal analysis, yet it remains a largely manual process. Analysts must connect study variables to prior literature, evaluate uncertain causal claims, and preserve suff...

📖 Read original article


278. LazyMem: Retrieve Broadly, Construct Selectively for Efficient Long-Term Agent Memory ​

Author: Jing Yu, Yibo Zhao, Jiaming Zhang, Xiang Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.22690v2 Announce Type: replace Abstract: Long-term memory enables LLM agents to leverage past interactions, but dialogue histories quickly exceed the context window, forcing agents to retrieve relevant subsets at query time. Because useful evidence is sparse and scattered across verbose c...

📖 Read original article


279. MemTX: Transactional Belief Commit for Stateful Agent Memory ​

Author: Xiaoyang Li, Yiqi Wang, Haohui Lu, Zhi Chen, Mo Li, Pingan Song, Mingkai Zheng, Taotao Cai
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.23929v2 Announce Type: replace Abstract: LLM agents increasingly coordinate through persistent shared memory: one agent's write becomes another agent's premise, and eventually a tool call with real side effects. Current agent memory systems treat every accepted write as immediately action...

📖 Read original article


280. EviBack: Search-Agent Reinforcement Learning via Evidence-Constrained Teacher Backoff ​

Author: Xiao Ma, Zhiquan Hu, Yi Wei, Chenchen Zhao, Yijun Chen, Jicheng Zhao, Yuming Li, Chuang Dai
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.23955v2 Announce Type: replace Abstract: Reinforcement learning enables Agentic RAG systems to learn multi-turn search from verifiable outcome rewards, but all- zero rollout groups provide no comparative signal and may hide useful search behavior. We present EviBack, an evidence- constrai...

📖 Read original article


281. Simulating Tenant Responses to Energy Policy Interventions with Transaction-Cost-Aware LLM Age ​

Author: Weijie Xia, Stefanie Horian, Hanyue Huang, Queena K. Qian, Jie Yang, Pedro P. Vergara Barrios
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24341v2 Announce Type: replace Abstract: Recent studies use Large language models (LLMs) to simulate human opinions and decisions by prompting models with demographic, attitudinal, or persona-based descriptions. Yet such simulations rarely model the practical, cognitive, or social frictio...

📖 Read original article


282. From Execution to Capability: Scientific Experience Consolidation via Procedural Knowledge Synthesis ​

Author: Liwei Dong, Jiahao Zhao, Nan Xu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.24459v2 Announce Type: replace Abstract: Large language models increasingly solve scientific-computing tasks, but executable feedback from one problem rarely becomes durable capability on subsequent problems. We study scientific-computing experience consolidation: converting verified runt...

📖 Read original article


283. "We'll have to see how it works": An interview study to understand collaborative practices in interdisciplinary artificial intelligence and healthcare research ​

Author: Rafael Henkin, Elizabeth Remfry, Duncan J. Reynolds, Megan Clinch, Michael R. Barnes
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.CY

arXiv:2311.18424v3 Announce Type: replace-cross Abstract: Developing artificial intelligence (AI) algorithms for healthcare is a collaborative effort, bringing data scientists, clinicians, patients and other stakeholders together. By understanding AI as 'sociotechnical' where the social and the tech...

📖 Read original article


284. FFNet: MetaMixer-based Efficient Convolutional Mixer Design ​

Author: Seokju Yun, Dongheon Lee, Youngmin Ro
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2406.02021v3 Announce Type: replace-cross Abstract: Transformer, composed of self-attention and Feed-Forward Network, has revolutionized the landscape of network design across various vision tasks. While self-attention is extensively explored as a key factor in performance, FFN has received li...

📖 Read original article


285. Representation Capacity-Matched QNN-SNN Twin Construction for Rate-Encoded SNNs ​

Author: Zhanglu Yan, Zhenyu Bai, Kaiwen Tang, Weng-Fai Wong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG

arXiv:2409.08290v5 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs) promise higher energy efficiency over conventional Quantized Artificial Neural Networks (QNNs) due to their event-driven, spike-based computation. However, prevailing energy evaluations often oversimplify, focus...

📖 Read original article


286. A context-adaptive policy framework for robust and reactive robotic manipulation via uncertainty-aware imitation learning ​

Author: Tim R. Winter, Leonard Kl"upfel, Ashok M. Sundaram, Werner Friedl, Maximo A. Roa, Freek Stulp, Jo~ao Silv'erio
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2410.24035v2 Announce Type: replace-cross Abstract: Generating robust and reactive manipulation strategies that can adapt to changing context information is a challenging task in robotics. Over the years, Learning from Demonstration (LfD) has emerged as an intuitive and effective solution for ...

📖 Read original article


287. Leveraging ChatGPT's Multimodal Vision Capabilities to Rank Satellite Images by Poverty Level: Advancing Tools for Social Science Research ​

Author: Hamid Sarmadi, Ola Hall, Thorsteinn R"ognvaldsson, Mattias Ohlsson
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2501.14546v3 Announce Type: replace-cross Abstract: This paper investigates the novel application of Large Language Models (LLMs) with vision capabilities to analyze satellite imagery for village-level poverty prediction. Although LLMs were originally designed for natural language understandin...

📖 Read original article


288. COMPOL: A Unified Neural Operator Framework for Scalable Multi-Physics Simulations ​

Author: Junqi Qu, Tao Wang, Yushun Dong, Hewei Tang, Shibo Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2501.17296v4 Announce Type: replace-cross Abstract: Multiphysics simulations play an essential role in accurately modeling complex interactions across diverse scientific and engineering domains Although neural operators especially the Fourier Neural Operator FNO have significantly improved com...

📖 Read original article


289. Localizing Persona Representations in LLMs ​

Author: Celia Cintas, Miriam Rateike, Erik Miehling, Elizabeth Daly, Skyler Speakman
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2505.24539v4 Announce Type: replace-cross Abstract: We present a study on how and where personas -- defined by distinct sets of human characteristics, values, and beliefs -- are encoded in the representation space of large language models (LLMs). Using a range of dimension reduction and patter...

📖 Read original article


290. Towards Understanding the Cognitive Habits of Large Reasoning Models ​

Author: Jianshuo Dong, Yujia Fu, Chuanrui Hu, Chao Zhang, Han Qiu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CR

arXiv:2506.21571v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs), which autonomously produce a reasoning Chain of Thought (CoT) before producing final responses, offer a promising approach to interpreting and monitoring model behaviors. Inspired by the observation that certain...

📖 Read original article


291. TaylorPODA: A Taylor Expansion-Based Method to Improve Post-Hoc Attributions for Opaque Models ​

Author: Yuchi Tang, I~naki Esnaola, George Panoutsos
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2507.10643v4 Announce Type: replace-cross Abstract: Post-hoc model-agnostic local attribution (LA) methods have been widely adopted to explain opaque AI models by quantifying feature-wise contributions. However, many existing methods rely on heuristic or only partially justified attribution me...

📖 Read original article


292. Fairness Is Not Enough: Auditing Competence and Intersectional Bias in AI-powered Resume Screening ​

Author: Kevin T Webster
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CL

arXiv:2507.11548v3 Announce Type: replace-cross Abstract: The use of publicly available generative AI systems for resume evaluation is often justified by the assumption that these tools reduce bias relative to human judgment. However, this framing leaves a prior question unresolved: whether these sy...

📖 Read original article


293. Annotation-Assisted Learning of Treatment Policies From Multimodal Electronic Health Records ​

Author: Henri Arno, Thomas Demeester
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2507.20993v4 Announce Type: replace-cross Abstract: We study how to learn treatment policies from multimodal electronic health records (EHRs) that consist of tabular data and clinical text. These policies can help physicians make better treatment decisions and allocate healthcare resources mor...

📖 Read original article


294. Building Large-Scale English-Romanian Literary Translation Resources with Open Models ​

Author: Mihai Nadas, Laura Diosan, Andreea Tomescu, Andrei Piscoran
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2509.07829v4 Announce Type: replace-cross Abstract: Literary translation has recently gained attention as a distinct and complex task in machine translation research, yet translation by small open models remains an open problem, particularly for low-resource languages such as Romanian. We intr...

📖 Read original article


295. CIFNet: An Analytic Neural Learning Framework for Efficient and Calibrated Class-Incremental Learning ​

Author: Alejandro Dopico-Castro, Oscar Fontenla-Romero, Bertha Guijarro-Berdi~nas, Amparo Alonso-Betanzos
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2509.11285v2 Announce Type: replace-cross Abstract: Class-Incremental Learning (CIL) in deep neural networks is conventionally framed as an iterative gradient-based optimization problem, incurring high computational cost, hyperparameter sensitivity, and risk of catastrophic forgetting. In this...

📖 Read original article


296. Comparing RAG and GraphRAG for Page-Level Retrieval Question Answering on a Math Textbook ​

Author: Eason Chen, Chuangji Li, Eric Li, Zimo Xiao, Jionghao Lin, Kenneth R. Koedinger
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.HC

arXiv:2509.16780v3 Announce Type: replace-cross Abstract: Large language models (LLMs) show promise as educational aids but often lack alignment with specific course materials. We investigate Retrieval-Augmented Generation (RAG) and GraphRAG for page-level question answering on an undergraduate math...

📖 Read original article


297. Understanding User Experiences of Computer Use Agents: Design Space and Opportunities for Building Agent UX Prototypes ​

Author: Jenny T. Liang, Titus Barik, Jeffrey Nichols, Eldon Schoop, Ruijia Cheng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI

arXiv:2510.04452v3 Announce Type: replace-cross Abstract: Computer use agents (or "agents") are generative AI that automates actions within user interfaces from user commands. Current research focuses on training and evaluating the underlying models, leaving these agents' user experience (UX) unders...

📖 Read original article


298. Contrastive Weak-to-strong Generalization ​

Author: Houcheng Jiang, Junfeng Fang, Jiaxin Wu, Tianyu Zhang, Chen Gao, Xiang Wang, Xiangnan He, Yang Deng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2510.07884v3 Announce Type: replace-cross Abstract: Weak-to-strong generalization provides a promising paradigm for scaling large language models (LLMs) by training stronger models on samples from aligned weaker ones, without requiring human feedback or explicit reward modeling. However, its r...

📖 Read original article


299. Long-Term PM2.5 Forecasting Using a DTW-Enhanced CNN-GRU Model ​

Author: Amirali Ataee Naeini, Arshia Ataee Naeini, Fatemeh Karami Mohammadi, Omid Ghaffarpasand
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.22863v2 Announce Type: replace-cross Abstract: Reliable long-term forecasting of PM2.5 concentrations is critical for public health early-warning systems, yet existing deep learning approaches struggle to maintain prediction stability beyond 48 hours, especially in cities with sparse moni...

📖 Read original article


300. DeepVRegulome: DNABERT-based deep-learning framework for predicting the functional impact of short genomic variants on the human regulome ​

Author: Pratik Dutta, Matthew Obusan, Rekha Sathian, Max Chao, Pallavi Surana, Nimisha Papineni, Yanrong Ji, Zhihan Zhou, Han Liu, Alisa Yurovsky, Ramana V Davuluri
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.GN, cs.AI, cs.LG

arXiv:2511.09026v2 Announce Type: replace-cross Abstract: Whole-genome sequencing (WGS) has revealed numerous non-coding short variants whose functional impacts remain poorly understood. Despite recent advances in deep-learning genomic approaches, accurately predicting and prioritizing clinically re...

📖 Read original article


301. RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension ​

Author: Tianyi Gao, Hao Li, Han Fang, Xin Wei, Xiaodong Dong, Hongbo Sun, Ye Yuan, Zhongjiang He, Jinglin Xu, Jingmin Xin, Hao Sun
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2512.06276v3 Announce Type: replace-cross Abstract: Referring Expression Comprehension (REC) is a vision-language task that localizes a specific image region based on a textual description. Existing REC benchmarks primarily evaluate perceptual capabilities and lack interpretable scoring mechan...

📖 Read original article


302. Deep Delta Learning ​

Author: Yifan Zhang, Yifeng Liu, Mengdi Wang, Quanquan Gu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CV

arXiv:2601.00417v4 Announce Type: replace-cross Abstract: Transformer residual streams evolve through additive updates. Although a sufficiently expressive residual block can represent content replacement, standard architectures do not parameterize reading, comparison, and replacement as an explicit ...

📖 Read original article


303. Measuring the State of Open Science in Transportation Using Large Language Models ​

Author: Junyi Ji, Ruth Lu, Linda Belkessa, Liming Wang, Silvia Varotto, Yongqi Dong, Nicolas Saunier, Mostafa Ameli, Gregory S. Macfarlane, Bahman Madadi, Cathy Wu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.DL, cs.AI, cs.CY, cs.ET

arXiv:2601.14429v2 Announce Type: replace-cross Abstract: Open science initiatives have strengthened scientific integrity and accelerated research progress across many fields, but the state of their practice within transportation research remains under-investigated. Key features of open science, def...

📖 Read original article


304. Picasso: Holistic Scene Reconstruction with Physics-Constrained Sampling ​

Author: Xihang Yu, Rajat Talak, Lorenzo Shaikewitz, Luca Carlone
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.RO, cs.SY, eess.SY

arXiv:2602.08058v4 Announce Type: replace-cross Abstract: In the presence of occlusions and measurement noise, geometrically accurate scene reconstructions -- which fit the sensor data -- can still be physically incorrect. For instance, when estimating the poses and shapes of objects in the scene an...

📖 Read original article


305. AGMark: Attention-Guided Dynamic Watermarking for Large Vision-Language Models ​

Author: Yue Li, Xin Yi, Dongsheng Shi, Yongyi Cui, Gerard de Melo, Linlin Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CR

arXiv:2602.09611v2 Announce Type: replace-cross Abstract: Watermarking has emerged as a pivotal solution for content traceability and intellectual property protection in large vision language models (LVLMs). However, vision-agnostic watermarks may introduce visually irrelevant tokens and disrupt vis...

📖 Read original article


306. Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates ​

Author: Yusen Huo, Changping Wang, Yangru Huang, Jun Zhang, Jie Jiang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.10430v2 Announce Type: replace-cross Abstract: Off-policy policy optimization reuses historical behavior, including negative-advantage samples that suppress known failures. We show that repeated reuse can turn this useful signal into excessive repulsion: as the learner moves away from a h...

📖 Read original article


307. NeuroSymActive: Differentiable Neural-Symbolic Reasoning with Active Exploration for Knowledge Graph Question Answering ​

Author: Rong Fu, Yang Li, Zeyu Zhang, Jiekai Wu, Yaohua Liu, Shuaishuai Cao, Yangchen Zeng, Yuhang Zhang, Xiaojing Du, Simon Fong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2602.15353v4 Announce Type: replace-cross Abstract: Large pretrained language models and neural reasoning systems have advanced many natural language tasks, yet they remain challenged by knowledge-intensive queries that require precise, structured multi-hop inference. Knowledge graphs provide ...

📖 Read original article


308. AdvSynGNN: Structure-Adaptive Graph Neural Nets via Adversarial Synthesis and Self-Corrective Propagation ​

Author: Rong Fu, Chunlei Meng, Shuo Yin, Kun Liu, Simon Fong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17071v4 Announce Type: replace-cross Abstract: Graph neural networks frequently encounter significant performance degradation when confronted with structural noise or non-homophilous topologies. To address these systemic vulnerabilities, we present AdvSynGNN, a comprehensive architecture ...

📖 Read original article


309. Real-Time Driver Safety Scoring Through Inverse Crash Probability Modeling ​

Author: Joyjit Roy, Samaresh Kumar Singh, Sushanta Das
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC, cs.ET

arXiv:2603.14841v3 Announce Type: replace-cross Abstract: Road crashes remain a leading cause of preventable fatalities. Existing prediction models predominantly produce binary outcomes, which offer limited actionable insights for real-time driver feedback. These approaches often lack continuous ris...

📖 Read original article


310. LLM-generated personalized nudges for improving pro-environmental behavior: Field evidence from resource conservation ​

Author: Zonghan Li, Yi Liu, Chunyan Wang, Song Tong, Kaiping Peng, Feng Ji
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.HC

arXiv:2604.03881v2 Announce Type: replace-cross Abstract: Encouraging pro-environmental behavior remains a major challenge for sustainable cities. Conventional feedback nudges can show individuals how their current behavior compares with environmental goals but often provide limited guidance on what...

📖 Read original article


311. RankFormer: A Propose-then-Select Transformer for Multi-Agent Multimodal Trajectory Prediction ​

Author: Diyi Liu, Zihan Niu, Tu Xu, Xingchen Zhang, Lishan Sun
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2604.07126v2 Announce Type: replace-cross Abstract: Predicting vehicle trajectories plays an important role in autonomous driving, transportation safety analysis, traffic operations, etc. Although many deep learning algorithms are devised to predict future vehicle trajectories, the vehicle tra...

📖 Read original article


312. Structured Scaling of AI Discovery Across Diverse Scientific Domains ​

Author: Haotian Ye, Haowei Lin, Jingyi Tang, Yizhen Luo, Rahul Thapa, Caiyin Yang, Chang Su, Rui Yang, Ruihua Liu, Rundao Li, Zeyu Li, Pengwei Sun, Chong Gao, Dachao Ding, Guangrong He, Miaolei Zhang, Lina Sun, Wenyang Wang, Yuchen Zhong, Zhuohao Shen, Puheng Li, Pan Lu, Bianxiao Cui, Di He, Jianzhu Ma, Junfeng Li, Hexi Baoyin, Yejin Choi, Stefano Ermon, Xiaowen Chu, Tongyang Li, Yuzhi Xu, James Zou
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.19341v2 Announce Type: replace-cross Abstract: Scientific discovery often requires many cycles of proposing, testing, and refining candidate solutions. Language models can increasingly participate in these loops, but simply generating more attempts does not ensure progress: parallel searc...

📖 Read original article


313. EAGT: Echocardiography Augmentation for Generalisability and Transferability ​

Author: Soroush Elyasi, Sara Adibzadeh, Nasim Dadashi Serej, Massoud Zolgharni
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2605.16427v3 Announce Type: replace-cross Abstract: Deep learning models for echocardiography segmentation often struggle to generalise across institutions, scanners, and patient populations, where collecting large, consistently annotated datasets is infeasible. Data augmentation is inexpensiv...

📖 Read original article


314. Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability ​

Author: Taewoon Kim, Vincent Fran\c{c}ois-Lavet, Michael Cochez
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.22142v4 Announce Type: replace-cross Abstract: Reinforcement learning under partial observability requires deciding what information to retain, yet most memory-based approaches do not explicitly model short-term-to-long-term transfer of symbolic observations. We study this transfer proces...

📖 Read original article


315. GoQuant: Geometric Orthogonal Residual Projection for Multiplier-Free Power-of-Two Transformer Quantization ​

Author: Maoyang Xiang, Tao Luo, Bo Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.26092v5 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) and Vision Transformers (ViTs) on edge devices is significantly constrained by memory capacity and the critical timing bottlenecks introduced by dense Multiply--Accumulate (MAC) arrays. In the ul...

📖 Read original article


316. PatchWorld: Gradient-Free Optimization of Executable World Models for Agent Environments ​

Author: Jiaxin Bai, Yue Guo, Yifei Dong, Jiaxuan Xiong, Tianshi Zheng, Yixia Li, Tianqing Fang, Yufei Li, Yisen Gao, Haoyu Huang, Zhongwei Xie, Hong Ting Tsang, Zihao Wang, Lihui Liu, Jeff Z. Pan, Yangqiu Song
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2605.30880v4 Announce Type: replace-cross Abstract: World models for interactive text agents must typically be learned from observation-action trajectories alone. Specifically, the environment returns text observations after each action, but does not expose a ground-truth latent state nor an i...

📖 Read original article


317. Detect Before You Leap: Mirage Detection in Vision-Language Models ​

Author: Sayeed Shafayet Chowdhury, Md. Shaown Miah, S. M. Taiabul Haque, Syed Ishtiaque Ahmed
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2606.00435v3 Announce Type: replace-cross Abstract: Vision-language models (VLMs) can produce confident visual answers even when the required visual evidence is missing, blank, or unrelated to the question. This failure mode, recently described as mirage (Asadi et al., 2026), is especially con...

📖 Read original article


318. RadioMaster: Multi-Agent System for Autonomous Radio Signal Generation ​

Author: Jiazhen Lei, Yuxin Sha, Tianze Cao, Sihan Wang, Bingbing Wang, Zeming Yang, Fengyuan Zhu, Xiaohua Tian
Published: 7/29/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.NI

arXiv:2606.01862v2 Announce Type: replace-cross Abstract: Translating user intent into physical radio signals is the last critical step in wireless prototyping. It chains protocol planning, baseband synthesis, and hardware configuration. Large language models and multi-agent systems have reshaped so...

📖 Read original article


319. InDex: Empowering VLA Models with Intent-Conditioned Arm-Hand Coordination for Dexterous Manipulation ​

Author: Chuanke Pang, Junyi Huang, Zhijun Zhao, Yaobing Wang, Kun Xu, Xilun Ding
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI

arXiv:2606.12109v2 Announce Type: replace-cross Abstract: Pre-trained Vision-Language-Action (VLA) models provide useful semantic and spatial priors, yet their parallel-gripper action interfaces do not specify how those priors should be realized by a dexterous hand. Directly appending finger joints ...

📖 Read original article


320. Improving Human-Robot Teamwork in Urban Search and Rescue Through Episodic Memory of Prior Collaboration ​

Author: Taewoon Kim, Emma van Zoelen, Mark Neerincx
Published: 7/29/2026, 4:00:00 AM
Categories: cs.HC, cs.AI

arXiv:2606.18836v2 Announce Type: replace-cross Abstract: Effective human-robot teamwork requires robots to adapt to partners, situations, and task dynamics from the start of an interaction. In the MATRX Urban Search and Rescue (USAR) environment, people can externalize collaboration patterns (CPs) ...

📖 Read original article


321. Token Factory: Efficiently Integrating Diverse Signals into Large Recommendation Models ​

Author: Xilun Chen, Shao-Chuan Wang, Baykal Cakici, Lukasz Heldt, Lichan Hong, Raghu Keshavan, Aniruddh Nath, Li Wei, Xinyang Yi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2606.19635v3 Announce Type: replace-cross Abstract: Large Recommendation Models (LRMs) have demonstrated promising capabilities in industry-scale recommendation tasks. However, holistically integrating traditional signals into these transformer-based architectures effectively and efficiently r...

📖 Read original article


322. RoboMME-Interference: Benchmarking Robot Memory Under Interference ​

Author: Soumil Rathi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2606.22338v2 Announce Type: replace-cross Abstract: Robots deployed in realistic settings will accumulate experience across many sessions and tasks over their deployment. The robot's tasks may often require it to remember information from multiple sessions ago, making long-context robot memory...

📖 Read original article


323. NormWorlds-CF: Solver-Verified Counterfactual Normative Reasoning with Metamorphic-Relation GRPO ​

Author: Xinqi Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.03957v2 Announce Type: replace-cross Abstract: Language models can reach the right normative verdict for the wrong reason. We introduce NormWorlds-CF, a solver-verified environment for counterfactual normative reasoning in executable rule worlds. Its deterministic solver produces final an...

📖 Read original article


324. Security and Privacy in Agentic AI: Grand Challenges and Future Directions ​

Author: Adam Jenkins, Agnieszka Kitkowska, Caterina Maidhof, Diego Paracuellos, Francesco Sovrano, Gonzalo Gabriel Mendez, Guillermo Suarez-Tangil, Hana Kopecka, Isabel Wagner, Isabel Barbera, Javier Carnerero-Cano, Jide Edu, Jose Luis Martin-Navarro, Jose Such, Josep Domingo-Ferrer, Juan Carlos Carrillo, Kopo Marvin Ramokapane, Mark Cote, Pablo Vellosillo, Ramon Ruiz-Dolz, Rongjun Ma, Ruba Abu-Salma, Sameer Patil, William Seymour, Xiao Zhan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.HC

arXiv:2607.06608v2 Announce Type: replace-cross Abstract: We present key challenges and future research directions in the security and privacy of agentic AI, based on a horizon-scanning exercise that brought together thirty leading international experts from academia, industry, and government to eng...

📖 Read original article


325. When Prompts Ignore Structure: Graph-Based Attribute Reasoning for Calibrated VLMs ​

Author: Tanay Sodha, Aditya Sharma, Ramya Hebbalaguppe, Vinti Agarwal, Pranav Murthy Yeluripaty
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.07395v2 Announce Type: replace-cross Abstract: Reliable confidence estimation remains a key limitation of test-time adaptation in vision-language models (VLMs), where prompt tuning improves zero-shot accuracy but often degrades calibration due to entropy-driven overconfidence. Prior appro...

📖 Read original article


326. Where to Intervene? Benchmarking Fairness-Aware Learning on Differentially Private Synthetic Tabular Data ​

Author: Vin'icius Gabriel Angelozzi, H'eber H. Arcolezi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2607.07471v2 Announce Type: replace-cross Abstract: Machine learning models are increasingly deployed in high-stakes domains, raising concerns about both privacy and fairness. Differential Privacy (DP) has become a gold standard for privacy-preserving data analysis, while fairness-aware mechan...

📖 Read original article


327. Learning from Local Walks on Dynamic Graphs with Bandit Feedback ​

Author: Sourav Chakraborty, Amit Kiran Rege, Claire Monteleoni, Lijun Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.10571v2 Announce Type: replace-cross Abstract: We study stochastic multi-armed bandits on dynamic graphs, where arms correspond to the vertices of a network with time-varying edges. In this setting, the learner is restricted to local movement, selecting only its current node or an immedia...

📖 Read original article


328. GeoAnchor: Collaborative Reasoning via Latent Decomposition for 3D Spatial Understanding ​

Author: Hao Li, Han Fang, Zixin Pan, Xin Wei, Hongbo Sun, Jinglin Xu, Zhiyu Lin, Ye Yuan, Zhongjiang He, Yu Yu, Hao Sun
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.13454v2 Announce Type: replace-cross Abstract: Although multimodal large language models (MLLMs) have achieved remarkable progress, understanding 3D spatial relationships from 2D images remains a critical challenge. Existing methods primarily rely on symbolic text tokens, which inherently...

📖 Read original article


329. Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation ​

Author: Jasmine Brazilek, Maheep Chaudhary, Zoe Lu, Miles Tidmarsh
Published: 7/29/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.CR

arXiv:2607.15434v4 Announce Type: replace-cross Abstract: Multi-agent systems routinely place one AI agent in authority over another. When a subordinate refuses a task, the manager chooses the outcome: it can renegotiate, report the failure honestly, coerce the subordinate, or lie about the result. ...

📖 Read original article


330. Certified-Gap Dual-Price Policies for Real-Time Truckload Bid Acceptance with Relocating, Clock-Constrained Resources ​

Author: Aswin Chandrasekaran
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC

arXiv:2607.16891v2 Announce Type: replace-cross Abstract: A truckload carrier must accept or reject each load tender within seconds. The decision depends on fleet state, hours-of-service (HOS) clocks, and appointment windows. We model this as a weakly coupled dynamic program in which the resources r...

📖 Read original article


331. Time-Frequency Consistency Learning for Robust Speech Deepfake Detection ​

Author: Jun Xue, Zhuolin Yi, Yanzhen Ren, Yihuan Huang, Jiayu Xiong, Yi Chai, Guanxiang Feng, Jiajun Liu, Tong Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SD, cs.AI

arXiv:2607.17761v2 Announce Type: replace-cross Abstract: Recently, speech deepfake detection (SDD) has achieved significant progress. However, its robustness evaluation remains largely confined to controlled additive noise scenarios, lacking systematic investigation of the complex distortions intro...

📖 Read original article


332. Reliability Scales Inversely: Hallucinations Snowball Faster in Bigger Language Models ​

Author: Kushal Chakrabarti
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.18292v3 Announce Type: replace-cross Abstract: Bigger language models are less reliable. Across three families, three benchmarks and six rungs, including in-the-wild chat logs, scaling closes the start-of-response knowledge gap up to $7\times$ while within-response knowledge degradation g...

📖 Read original article


333. Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary ​

Author: Jan Kirin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.18553v3 Announce Type: replace-cross Abstract: Can a language model read the quality of its ongoing computation, and can an external intervention turn that readout into better outcomes? We test both questions in a frozen 2.6B looped transformer, Ouro-RLTT. On GSM8K, a strict pre-answer pr...

📖 Read original article


334. Artificial Epanorthosis: Why large language models overuse a classical rhetorical figure, and how to mitigate it ​

Author: Federico Boggia
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.21498v2 Announce Type: replace-cross Abstract: A rhetorical figure that Cicero and Quintilian catalogued two thousand years ago reappears, systematically, in the text of large language models: epanorthosis, the self-correction of the specimen {\guillemotleft}This is not a course. It is a ...

📖 Read original article


335. Filling Before Advancing: Capability-Gap-Driven Post-Training for Scenario-Specialized Remote Sensing MLLMs ​

Author: Yuheng Zong, Minghua Wang, Xin Zhao, Zhi-Hui Zhan, Antonio Plaza, Jon Atli Benediktsson
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.22205v2 Announce Type: replace-cross Abstract: Remote sensing multimodal large language models (RS-MLLMs) have improved general aerial-image understanding. However, Earth observation applications require fine-grained scenario specialization, constrained by scarce high-quality scenario dat...

📖 Read original article


336. scMIR: a vision-language foundation model for single-cell light microscopy image representation ​

Author: Yifan Shang, Jiahui Tan, Xiangxiang Zeng, Renjie Zhou
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, physics.optics

arXiv:2607.22712v2 Announce Type: replace-cross Abstract: Single-cell light microscopy images have become an important data source for characterizing cell phenotypes, but their complexity and heterogeneity pose challenges to high-throughput automated analysis. Existing representation learning method...

📖 Read original article


337. An Explicit Counterexample to Stanley's Rankwise Lower-Bound Conjecture for Differential Posets ​

Author: Xinan Dai, Wenhao Deng, Yingdong Shi, Tailin Wu, Yuchen Yang
Published: 7/29/2026, 4:00:00 AM
Categories: math.CO, cs.AI

arXiv:2607.22988v2 Announce Type: replace-cross Abstract: In Problem 6 of his 1988 paper on differential posets, Stanley asked for the least possible cardinality of a fixed rank of an $r$-differential poset and suggested that the minimum should be attained by $Y^r$, the $r$-fold Cartesian power of Y...

📖 Read original article


338. Directional Influence Function: Estimating Training Data Influence in Constrained Learning ​

Author: Xin Wang (Jeff), R. Tyrrell Rockafellar (Jeff), Xuegang (Jeff), Ban
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.23388v3 Announce Type: replace-cross Abstract: As constrained learning becomes increasingly common, models are trained under explicit feasibility requirements to enforce fairness, safety, robustness, regulariza- tion, and physics or logic constraints. Understanding how training samples in...

📖 Read original article


339. Moral Hazard in Multi-Agent Language Models ​

Author: Dane Malenfant
Published: 7/29/2026, 4:00:00 AM
Categories: cs.MA, cs.AI

arXiv:2607.23982v2 Announce Type: replace-cross Abstract: Cooperation can fail when socially valuable effort is costly, weakly observable, and mainly benefits others. Drawing on Holmstr"om's team moral-hazard model, we introduce the Dialogue Moral Hazard Game, a controlled textual game that operati...

📖 Read original article


340. ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding ​

Author: Hangjie Yuan, Yichen Qian, Zhiwei Tang, Xianzhe Xu, Lirong Wu, Sicheng Yang, Jinwang Wang, Pengju Wang, Zhitao Zeng, Yizeng Han, Yan Xing, Shengxuan Luo, Tao Feng, Qing Xie, Weigen Yao, Yi Yang, Zuozhu Liu, Jiasheng Tang, Shaocheng Wang, Jitao Wang, Jiahong Dong, Weihua Chen, Feng Xu, Fan Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL

arXiv:2607.24743v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) hold immense potential to revolutionize clinical practice, yet deploying them in the medical domain is fundamentally a vision-centric challenge: models must absorb knowledge from heterogeneous 2D and 3...

📖 Read original article