Skip to content

arXiv cs.AI - 2026-07-11 ​

253 items collected.


1. Context Graphs for Proactive Enterprise Agents ​

Author: Avinash Kumar
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.07721v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) and agentic frameworks have advanced enterprise AI considerably, yet agents remain fundamentally reactive: they wait for a human query before acting. This paper argues that genuine enterprise productivity gains requ...

📖 Read original article


2. AI-integrated models for assessing agricultural resilience ​

Author: Joshua R. Waite, Dana Golden, Brett Indelicato, Kevin Camp, Mojdeh Saadati, Shannon Regan, Patrick Schnable, Baskar Ganapathysubramanian, Carlos Messina, Suzanne Thornsbury, Soumik Sarkar
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.07759v1 Announce Type: new Abstract: Agricultural supply chains are vulnerable to disruptions through linked biophysical and economic systems. We develop an AI-powered tool that integrates economic models (GTAP) with biophysical models (APSIM) to analyze supply chain shocks, enabling poli...

📖 Read original article


3. Adversarial Social Epistemology for Assemblies of Humans and Large Language Models ​

Author: Mihnea C. Moldoveanu, Joel A. C. Baum
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.SI

arXiv:2607.07760v1 Announce Type: new Abstract: We outline an adversarial social epistemology (ASE) for densely interactive communicative landscapes in which public assertions are scaffolded by chains of testimony, inference, institutional certification, and tacit trust. In such landscapes, agents h...

📖 Read original article


4. Aligning Clinical Needs and AI Capabilities: A Survey on LLMs for Medical Reasoning ​

Author: Qi Peng, Jiatong Li, Sirui Huang, Yiyang Jiang, Kaisong Gong, Ronger Ding, Shijie Ye, Changmeng Zheng, Yi Cai, Xiaobo Yang, Jin Huang, Xiao-Yong Wei, Qing Li
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.07761v1 Announce Type: new Abstract: Large language models (LLMs) have emerged as important tools in healthcare, showing growing potential for clinical reasoning and patient care. This survey examines recent progress in medical LLMs, focusing on reasoning applications and requirements. We...

📖 Read original article


5. Alignment Plausibility: A New Standard for Assuring AI in Healthcare ​

Author: Gwydion Williams, Sara Zannone, Bilal A Mateen
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.07766v1 Announce Type: new Abstract: Large language models (LLMs) have become significant providers of mental health support, yet they remain products of an attention economy whose operational and commercial targets favour sustained engagement over the friction that effective psychologica...

📖 Read original article


6. Idiobionics: The Unification of Privacy and Intelligent Robotic Prostheses ​

Author: Kwesi Afari Darfoor, Patrick M. Pilarski, Bailey Kacsmar
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.HC, cs.RO

arXiv:2607.07775v1 Announce Type: new Abstract: The human body is at the center of a growing family of technologies designed to tightly and persistently couple biological and digital systems. Robotic prostheses are a representative example of this tight coupling. Also referred to as bionic limbs, ro...

📖 Read original article


7. Infinity-Parser2 Technical Report ​

Author: Zuming Huang, Jun Huang, Kexuan Ren, Baode Wang, Weizhen Li, Jianming Feng, Yu Wang, Yichen Yao, Shijun Lin, Yige Tang, Cheng Peng, Weidi Xu, Wei Chu, Yinghui Xu, Yuan Qi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.07836v1 Announce Type: new Abstract: We present Infinity-Parser2, a large multimodal model that couples a controllable data-synthesis pipeline with multi-task reinforcement learning for end-to-end document parsing, addressing the persistent scarcity of faithfully annotated parsing corpora...

📖 Read original article


8. VectorizationLLM: Smart Vectorization Based AI Assistant ​

Author: Ryan Duke
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CY

arXiv:2607.07846v1 Announce Type: new Abstract: VectorizationLLM is a specialized Large Language Model based on Google open-weight LLMs. The model is designed to assist students to learn smart vectorization, time/wave vector analysis, piecewise functions, Fourier analysis, and differential equations...

📖 Read original article


9. A Graph Neural Network Model for Real-Time Gesture Recognition Based on sEMG Signals ​

Author: Pragatheeswaran Vipulanandan, Kamal Premaratne, Manohar Murthi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.07850v1 Announce Type: new Abstract: For seemless control of advanced hand prostheses and augmented reality, accurate and immediate hand gestures recognition is essential. Surface electromyography (sEMG) signals obtained from the forearm are commonly employed for this purpose. In this pap...

📖 Read original article


10. Agentic AI and Retrieval-Augmented Models in Straight-Through Underwriting ​

Author: Robert Richardson, Josh Meyers, Brian Hartman, David Sandberg
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.07858v1 Announce Type: new Abstract: Artificial intelligence (AI) is beginning to reshape actuarial practice, particularly in domains that require reasoning over unstructured documents, heterogeneous data sources, and regulated decision workflows. Actuaries now face a design space that ra...

📖 Read original article


11. Feedback Manipulation Regularization: Enabling Offline Agent Alignment for Imitation Learning ​

Author: Benjamin Poole, Minwoo Lee
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.HC, cs.LG

arXiv:2607.07859v1 Announce Type: new Abstract: Reinforcement learning (RL) research has increasingly shifted focus towards alignment, ensuring agents learn behaviors adhering to human values. While human demonstrations and feedback have proven crucial for alignment, existing approaches predominantl...

📖 Read original article


12. Nigeria Machinery: A Low-Resource Industrial Dataset with a Domain-Grounded Reasoning Layer ​

Author: Gospel Bassey, Vincent Fakiyesi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.07883v1 Announce Type: new Abstract: There is relatively little, public, and model-ready data on industrial machinery for African economies. This makes it hard to do quantitative analysis or to train language models on numeric tasks grounded in that setting. We release two things to help ...

📖 Read original article


13. Persona Cartography: Charting Language Model Personality Traits in Weight Space ​

Author: Luke Baines, Anton Gonzalvez Hawthorne, Mariia Koroliuk, Irakli Shalibashvili, Cl'ement Dumas, Konstantinos Voudouris, David Demitri Africa
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.07916v1 Announce Type: new Abstract: Large language models exhibit recurring behavioural patterns -- personas -- that shape generalisation and safety, but we lack reliable tools for decomposing, measuring, and controlling them. Our central insight is to treat personas as positions in a sp...

📖 Read original article


Author: Raunak Mondal, Peter Washington
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG

arXiv:2607.07957v1 Announce Type: new Abstract: Autism spectrum disorder (ASD) affects over 75 million individuals worldwide, yet scalable computational methods for remote behavioral screening remain limited. This study addresses two complementary challenges in automated detection of autism-related ...

📖 Read original article


Author: Seokhoon Jeong, Mijung Kim, Taehwan Kim
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.07984v1 Announce Type: new Abstract: Neural architecture search (NAS) methods have grown increasingly efficient, yet they remain bounded by manually engineered search spaces that require substantial domain expertise and must be rebuilt for every new task. Large language models (LLMs) can ...

📖 Read original article


16. Concretized Proposition Prompting Resolves Composition-Knowledge Dichotomy in Large Language Models ​

Author: Changhun Lee, Minguk Jeon, Jongkyung Shin, Chiehyeon Lim
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08018v1 Announce Type: new Abstract: LLMs often struggle to balance compositionality with knowledgeability, a challenge we define as Composition-Knowledge Dichotomy. To address this, we propose Concretized Proposition Prompting (CPP), a framework that explicitly concretizes propositions r...

📖 Read original article


17. From Prompts to Contracts: Harness Engineering for Auditable Enterprise LLM Agents ​

Author: Joongho Ahn, Moonsoo Kim
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.SE

arXiv:2607.08028v1 Announce Type: new Abstract: Enterprise large language model (LLM) applications often begin as prototypes whose behavior is carried by prompts and retrieval context. Productization adds requirements for source boundaries, entity routing, answer contracts, and reproducible traces. ...

📖 Read original article


18. A safety-oriented hypothetico-deductive framework for AI-assisted differential diagnosis ​

Author: Fan Ma, Mauro Giuffr`e, Donald Wright, Kent McCann, Mark Iscoe, Lingfei Qian, Mingyang Jiang, Chi Wing Ng, Na Hong, Huan He, Cathy Shyr, Qingyu Chen, Lee Schwamm, Lucila Ohno-Machado, Hua Xu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08038v1 Announce Type: new Abstract: Diagnostic error is a major threat to patient safety, yet current large language model (LLM) systems often treat diagnosis as a one-shot prediction task, lacking safeguards against missed high-risk alternatives or rigorous verification of their reasoni...

📖 Read original article


19. When LLMs Agree, Are They Right? Auditing Self-Consistency and Cross-Model Agreement as Confidence Signals ​

Author: Kaihua Ding
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08065v1 Announce Type: new Abstract: LLM-as-judge (Zheng et al., 2023) is increasingly the default for evaluating AI systems in enterprise pipelines, often scaled to ensembles (Verga et al., 2024) or "mixture-of-experts" (Shazeer et al., 2017) panels of judges. These systems share a key a...

📖 Read original article


20. Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring ​

Author: Jennifer Za, Julija Bainiaksina, Nikita Ostrovsky, Tanush Chopra, Victoria Krakovna
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.08066v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is a promising safety mechanism for AI agents, based on the premise that visible reasoning traces can surface misaligned or deceptive behavior. While effective in standard scenarios, recent work highlights that LLMs re...

📖 Read original article


21. PARA-PV: Physics-Aware Retrieval-Augmented PV Prediction Based on Frozen Foundation Model and Distribution Shift Correction ​

Author: Hang Fan, Weican Liu, Ying Lu, Dunnan Liu, Long Cheng, Wei Wei
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08079v1 Announce Type: new Abstract: Accurate photovoltaic (PV) power forecasting is essential for reliable grid dispatch and renewable energy integration, yet it remains challenging because PV generation is jointly shaped by weather variability, day-night transitions, regime-dependent dy...

📖 Read original article


22. CausalDS: Benchmarking Causal Reasoning in Data-Science Agents ​

Author: Andrej Leban, Yuekai Sun
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.08093v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as integrated data-science agents, combining abstract reasoning with advanced tool use. Yet the relevant benchmark landscape largely divides into symbolic causal reasoning benchmarks without realistic data ...

📖 Read original article


23. Answer Set Programming Energised! End-to-End Neurosymbolic Reasoning and Learning with ASP and Energy Based Models ​

Author: Jakob Suchan, Julius Monsen, Salim Baloch, Mehul Bhatt
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08136v1 Announce Type: new Abstract: We present a general neurosymbolic reasoning and learning methodology based on a modular integration of answer set programming with an energy based model substrate. Key contributions are: (1) supporting joint optimisation in the continuous latent space...

📖 Read original article


24. Overthinking: Amplifying Reasoning Weights to Extract Learned Secrets ​

Author: Jack Hopkins, Dipika Khullar, Fabien Roger
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08173v1 Announce Type: new Abstract: Black box auditing of language models is an essential pre-deployment tool, but it may miss subtle forms of misalignment and hidden information. To better elicit hidden information during an auditing process, we introduce \emph{overthinking}: the proces...

📖 Read original article


25. ASMR: Agentic Schema Generation for Ship Maintenance Report Writing ​

Author: Sohrab Namazi Nia, Amogh Dalal, Ning Sa, Peter Ly, Marti Zentmaier, Tomek Strzalkowski, Jay Miller, Rishi Singh, Senjuti Basu Roy
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.MA

arXiv:2607.08177v1 Announce Type: new Abstract: In this paper, we study the automatic schema generation problem: given a collection of historical ship maintenance and operational reports across multiple form categories, automatically discover compact and informative schemas that capture the essentia...

📖 Read original article


26. A First-Principles Theory of Slow Thinking and Active Perception ​

Author: Hongkang Yang, Zhi-Qin John Xu, Feiyu Xiong, Weinan E
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.08196v1 Announce Type: new Abstract: As part of a series on first-principles modeling of cognitive functions, this paper attempts to provide a mathematical formulation of thinking and perception. It formally derives slow thinking or more generally, active perception, and encompasses the d...

📖 Read original article


27. Playing ZendoWorld: Challenging AI Agents on Active Visual Concept Induction ​

Author: Sophia Koehler, Antonia W"ust, Inga Ibs, Wasu Top Piriyakulkij, Wolfgang Stammer, Constantin Rothkopf, Kevin Ellis, Kristian Kersting
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CV

arXiv:2607.08233v1 Announce Type: new Abstract: A central challenge in building intelligent systems is enabling agents to jointly perceive complex inputs, form hypotheses about hidden patterns, and design informative experiments to test them. To study this problem, we propose ZendoWorld, a controlle...

📖 Read original article


28. AutoPersonas: A Multi-Timescale Loop Engine for Open-Ended Persona Evolution ​

Author: Mengchen Li
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.HC

arXiv:2607.08252v1 Announce Type: new Abstract: Long-term persona agents must remain identifiable while adapting to new events, relationships, evidence, and social conditions. We identify self-locking as a runtime failure mode in continuing persona-life loops: locally plausible events keep appearing...

📖 Read original article


29. Compete Then Collaborate: Frontier AI Teachers Build a Verifiable Curriculum to Improve a Coding Student Beyond Imitation ​

Author: Miseong Shawn Kim
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08255v1 Announce Type: new Abstract: Large language models increasingly serve as teachers generating training data for smaller students. Prior multi-teacher knowledge distillation methods merge outputs without determining which frontier model teaches best, often relying on an LLM judge bi...

📖 Read original article


30. MentalHospital: A Virtual Environment for Evaluating Psychiatric Clinical Encounters ​

Author: Yuming Yang, Xiao Sun, Yuanwei Zou, Zhengxiao Wu, Yun Chen, Jiang Zhong, Haoyang Zeng, Jingwang Huang, Kaiwen Wei
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08257v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong performance on isolated psychiatric tasks, including dialogue, diagnosis, and treatment planning, yet existing benchmarks rarely simulate complete psychiatric clinical encounters. We introduce $\textbf{Men...

📖 Read original article


31. Different Teachers, Different Capabilities: Sub-1B On-Device Distillation for Structured Text Enrichment ​

Author: Vinay Kumar Chaganti
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.08268v1 Announce Type: new Abstract: High-volume structured extraction pays a large model's latency on every item, so distilling the task into a small on-device model is attractive: comparable output at a fraction of the time and cost. We measure what that distillation actually delivers, ...

📖 Read original article


32. PolyUQuest: Verifiable Structure-Aware Web RAG over Heterogeneous Graphs ​

Author: Ying Liu, Yi Ye, Quanyu Feng, Mingxi Ye, Mingtao Zhang, Haoyang Li, Chen Jason Zhang, Qing Li
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08269v1 Announce Type: new Abstract: Existing retrieval-augmented generation (RAG) systems treat web pages as flat text, losing the structural and semantic signals encoded in HTML. We present PolyUQuest, a verifiable, structure-aware web RAG framework built on a heterogeneous graph that u...

📖 Read original article


33. Understanding Axes of Difficulty For Long Context Tasks Via PredicateLongBench ​

Author: Siddhartha Jain, Ameya Velingker
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08284v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated rapidly improving long-context capabilities, prompting a wave of benchmarks designed to evaluate them. However, existing long-context evaluations - from Needle-in-a-Haystack (NIAH) tests to more recent mul...

📖 Read original article


34. Psychological Competence as a Missing Dimension in AI Evaluation ​

Author: Marcos Economides, Paul M. Sacher, Samuel Salzer, Alexis Michelle Abellar, Fendi Tsim, Antoine Ferr`ere
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08285v1 Announce Type: new Abstract: Current AI evaluation frameworks focus primarily on technical performance, including accuracy, robustness, reasoning ability, and policy compliance. These measures remain essential, but they are not sufficient for systems that interact directly with us...

📖 Read original article


35. INTENT: An LSTM Framework for Vehicle Intention Prediction in Intersection Scenarios with Comprehensive Ablation Analysis ​

Author: Logine M. Zaki, Catherine M. Elias
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.RO

arXiv:2607.08316v1 Announce Type: new Abstract: Vehicle intention prediction is a pivotal aspect in the agility and safety of autonomous vehicles in all driving scenarios; if genuine enhancement of autonomous vehicles are required, we need to make them adopt human interpretation of driver's intentio...

📖 Read original article


36. Blind-Spots-Bench: Evaluating Blind Spots in Multimodal Models ​

Author: Matteo Santelmo, Xiuying Wei, Israa Fakih, Felix Bauer, Juan Garcia Giraldo, Chengkun Li, Etienne Bamas, Emmanuel Abb'e
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08317v1 Announce Type: new Abstract: Modern AI models achieve strong performance on many established benchmarks, yet they still fail on tasks that humans find almost trivial, such as manipulating a string or drawing a dog with five legs. These examples suggest that existing benchmarks may...

📖 Read original article


37. MobiDiff: Semantic-Aware Multi-Channel Discrete Diffusion for Human Mobility Data Generation ​

Author: Rongchao Xu, Lin Jiang, Dahai Yu, Ximiao Li, Taichi Liu, Desheng Zhang, Yuan Tian, Guang Wang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08357v1 Announce Type: new Abstract: Human mobility data are essential for transportation optimization, urban planning, and resource allocation, yet real-world mobility data are costly to collect and difficult to share due to privacy concerns. Recent diffusion-based methods have shown pro...

📖 Read original article


38. FedOPAL: One-Shot Federated Learning via Analytic Visual Prompt Tuning ​

Author: Lingyu Qiu, Daniela Annunziata, Stefano Izzo, Fabio Giampaolo, Francesco Piccialli
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08368v1 Announce Type: new Abstract: With the widespread deployment of basic models in edge intelligence, communication bandwidth has become a core bottleneck restricting the scalability of federated learning. Although one-shot federated learning alleviates this problem by minimizing comm...

📖 Read original article


39. Towards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model Finetuning ​

Author: Lu Dai, Ziyang Rao, Yili Wang, Hanqing Wang, Hao Liu, Hui Xiong
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.08393v1 Announce Type: new Abstract: Fine-tuning LLMs to inject new knowledge faces a critical challenge: LLMs can quickly memorize new facts, yet fail to use them for downstream reasoning tasks. We formalize this failure as the \textit{\textbf{Knowing--Using Gap}}, characterized by an ac...

📖 Read original article


40. Game Theory Driven Multi-Agent Framework Mitigates Language Model Hallucination ​

Author: Runzhe Liu, Biquan Bie, Zihao Wang, Yuchao Ma, Yexin Liu, Xinghai Li, Harry Yang, Wenbo Yang, Jinzhe Cao, Shengyang Tao
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08403v1 Announce Type: new Abstract: The application of lightweight Large Language Models in rule-based scientific domains remains severely limited by their tendency to mimic linguistic patterns rather than reproduce axiomatic reasoning, causing frequent hallucinations. Here, we show that...

📖 Read original article


41. OmniFood-Bench: Evaluating VLMs for Nutrient Reasoning and Personalized Health Advice ​

Author: Qian Jiang, Zhecheng Shi, Jingpu Yang, Zirui Song, Miao Fang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08423v1 Announce Type: new Abstract: The rapid integration of Large Vision-Language Models (VLMs) into critical infrastructure promises to revolutionize personalized healthcare and dietary management. However, in the domain of food systems, autonomous agents face a unique and persistent c...

📖 Read original article


42. Applying JEPA-Style Predictive Learning to JA4-Derived Network Fingerprints ​

Author: Javier Izquierdo, Aygul Zagidullina
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08465v1 Announce Type: new Abstract: I-JEPA and V-JEPA learn by matching latent predictions to target encoder outputs rather than regenerating the original input, and this has worked well for images and video. We explore whether the same objective works for compact network fingerprints. W...

📖 Read original article


43. Drift-Aware Temporal Graph Rewiring (DATGR) for Adaptive Semantic Modeling in Biomedical Text ​

Author: Bharathwaj Vijayakumar, Sahana K. Varadaraju
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08490v1 Announce Type: new Abstract: Biomedical language evolves rapidly as new discoveries emerge, causing traditional text models to lose semantic fidelity over time. Static embeddings and co-occurrence graphs cannot capture such evolution, leading to performance degradation in retrieva...

📖 Read original article


44. AI-guided stimuli discovery and generation to optimize facial emotion perception studies in autism ​

Author: Kushin Mukherjee, Na Yeon Kim, Maren Wehrheim, Ralph Adolphs, Kohitij Kar
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.08533v1 Announce Type: new Abstract: Understanding perceptual differences between autistic and neurotypical adults requires behavioral assays that are sensitive, reliable, and mechanistically informative. Facial emotion perception is a useful test case because group differences have been ...

📖 Read original article


45. CommuniWave:A Machine Learning Model for Quantifying the Degree of Temporary Informal Behavior in Urban Communities ​

Author: Hongye Yang, Shien Liu, Zhihao Xie
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08554v1 Announce Type: new Abstract: For urban managers and designers, improving the functional attributes of urban communities to enhance territorial resilience in the face of complexity and uncertainty is crucial. Currently, community planning often follows a top-down approach and lacks...

📖 Read original article


46. SHAP-Weighted Cross-Modal Expert Fusion for Emotion and Sentiment Recognition: Evidence and Limits ​

Author: Adis Alihodzic, Selma Skopljakovic Hubljar
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08573v1 Announce Type: new Abstract: Multimodal emotion and sentiment recognition is commonly addressed by early fusion, which concatenates modalities before classification, or late fusion, which combines independently trained unimodal predictors. Early fusion can be accurate but monolith...

📖 Read original article


47. Towards Precision Therapy in Hepatocellular Carcinoma: A Clinical-Reasoning LLM for Risk Stratification and Treatment Guidance ​

Author: Peng Cui, Jitao Wang, Siyan Xue, Yao Huang, Haoming Xia, Dong Li, Dengxiang Liu, Weilin Wang, Liping Liu, Leida Zhang, Yunfu Cui, Tao Peng, Daolin Ji, Haitao Zhao, Wei Zhang, Xiaojuan Wang, Weijie Ma, Zongren Ding, Jinlong Li, Yuan Ding, Jiajing Zhao, Zhiyu Chen, Chengkun Yang, Ziyue Huang, Jiaqi Liu, Fusheng Liu, Yang Zhou, Xiaojuan Wang, Zhongquan Sun, Shiyun Bao, Xiaojun Wang, Ming Yang, Guangxin Li, Bin Shu, Yong Liao, Hongxuan Li, Yao Tang, Shizhong Yang, Yongyi Zeng, Yufeng Yuan, Yinpeng Dong, Jihui Hao, Jun Zhu, Jiahong Dong
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08602v1 Announce Type: new Abstract: Hepatocellular carcinoma (HCC) is a common malignancy and a leading cause of cancer-related mortality. Current guidelines and staging systems provide coarse categories, but often miss within-stage heterogeneity and the clinical context in electronic me...

📖 Read original article


48. The complexities of patient-centred conversational artificial intelligence ​

Author: Jo~ao Matos, Olivia Buege, Donny Cheung, Gary S. Collins, Paula Dhiman, Nan Li, Bingyu Mao, Benjamin W. Nelson, Michail Ouroutzoglou, Paul Varghese, Jonathan Amar
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.08625v1 Announce Type: new Abstract: Consumer-facing health chatbots powered by large language models (LLMs) are increasingly used for symptom assessment. However, chatbot development and evaluation often rely on cooperative, articulate, simulated patients. We analysed 2,053 real patient-...

📖 Read original article


49. Formal Mechanisms for Market Stability in Self-Interested Agent Societies: A Marketplace Simulation Study ​

Author: Eugene Ng Yi Sheng, Bingquan Shen
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08652v1 Announce Type: new Abstract: Self-interested agents, left unconstrained, tend toward defection in repeated social dilemmas, causing cooperative gains from trade to collapse. This paper investigates what formal mechanisms, layered on top of unrestricted communication, are sufficien...

📖 Read original article


50. SolarChain-Eval: A Physics-Constrained Benchmark for Trustworthy Economic Agents in Decentralized Energy Markets ​

Author: Shilin Ou, Yifan Xu, Luyao Zhang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08681v1 Announce Type: new Abstract: As agentic AI systems are increasingly applied to cyber-physical environments, their evaluation requires assessment of both task performance and trustworthiness. In decentralized energy markets, autonomous agents may improve market utility, but may als...

📖 Read original article


51. Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents ​

Author: Yifan Wu, Lizhu Zhang, Yuhang Zhou, Mingyi Wang, Bo Peng, Serena Li, Xiangjun Fan, Zhuokai Zhao
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.08716v1 Announce Type: new Abstract: In long-horizon tasks, decision-relevant state is often scattered across an expanding trajectory, while the action agent must surface it and act. As trajectories grow, task requirements, environment facts, prior attempts, diagnoses, and open subgoals c...

📖 Read original article


52. The Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMs ​

Author: Baha Rababah, Cuneyt Gurcan Akcora, Carson K. Leung
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08734v1 Announce Type: new Abstract: Post-training quantization is widely used to deploy large language models in resource-constrained settings, yet its evaluation relies almost exclusively on accuracy and perplexity. We show that these metrics fail to capture behavioral changes induced b...

📖 Read original article


53. Workflow as Knowledge: Semantic Persistence for LLM-Mediated Workflows ​

Author: Emanuele Quinto, Carlo Andrea Rozzi, Francesco Zanitti
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.PL, cs.SE

arXiv:2607.08740v1 Announce Type: new Abstract: Large language model (LLM) applications increasingly use explicit workflows for tool use, retrieval, branching, checkpointing, and human approval. Existing workflow systems already address many execution concerns. This paper proposes a Lisp-inspired bu...

📖 Read original article


54. AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding ​

Author: Siddharth Damodharan, Radhika Gupta, Ali Alshami, Ryan Rabinowitz, Jugal Kalita
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CV

arXiv:2607.08745v1 Announce Type: new Abstract: Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as scene understanding, decision making, trajectory prediction, and visual question answering. However, e...

📖 Read original article


55. Using AI-based Learning Assistants in Higher Education: A Large-Scale Descriptive Analysis ​

Author: Kristina Schaaff, Quintus Stierstorfer, Valerie Heckel
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.HC

arXiv:2607.08748v1 Announce Type: new Abstract: In this study, we present a large-scale descriptive analysis of the use of an AI-based learning assistant (Syntea) in higher education. Based on objective log data from 77,543 students enrolled in distance studies, we examine usage patterns across gend...

📖 Read original article


56. Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation ​

Author: Yifan Zhou, Qihao Yang, Yan Li, Donggang Li, Xiru Hu, Hokin Deng, Ziyang Gong, Xuanyi Zhou, Huacan Wang, Xiangchao Yan, Wanghan Xu, Wenlong Zhang, Shaofeng Zhang, Yue Zhou, Yifan Yang, Zhihang Zhong, Xue Yang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.08758v1 Announce Type: new Abstract: Scientific ideas rarely start from a blank page. They inherit mechanisms, repair known limitations, and recombine pieces of earlier work, much like biological genomes. Current benchmarks still say little about whether AI systems can follow this inherit...

📖 Read original article


57. Towards the Explainability of Temporal Graph Networks via Memory Backtracking and Topological Attribution ​

Author: Yazheng Liu, Xi Zhang, Sihong Xie, Hui Xiong
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.07716v1 Announce Type: cross Abstract: Temporal graphs are ubiquitous in real-world applications and Temporal Graph Networks (TGNs) have achieved superior predictive accuracy. Understanding which historical events drive model predictions can enhance trustworthiness of TGNs. Existing expla...

📖 Read original article


58. LLT: Local Linear Transformer for PDE Operator Learning ​

Author: Oded Ovadia, Eli Turkel
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NA, math.NA

arXiv:2607.07718v1 Announce Type: cross Abstract: Neural operators have become a common approach for learning PDE solution maps and accelerating numerical simulations. Transformer-based neural operators are of particular interest, since attention can learn long-range dependencies in the computationa...

📖 Read original article


59. ReCoLoRA: Spectrum-Aware Recursive Consolidation for Continual LLM Fine-Tuning ​

Author: Wentao Lu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.07719v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning adapts a large language model to one task cheaply, but across a task sequence LoRA-style methods keep stacking low-rank updates on the same frozen weight, so each new task tends to overwrite the previous ones. We prese...

📖 Read original article


60. Omni-Sleep: A Sleep Foundation Model via Hierarchical Contrastive Learning of CNS--ANS Dynamic ​

Author: Zhoujie Hou, Song Wang, Kexin Lou, Mo Wang, Chen Wei, Quanying Liu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.07720v1 Announce Type: cross Abstract: Sleep physiology arises from the coordinated dynamics of the central nervous system (CNS) and autonomic nervous system (ANS), as reflected by multimodal polysomnography signals including EEG, EOG, EMG, ECG, and respiration. However, existing sleep fo...

📖 Read original article


61. SHIFT: Survival Prediction from Incomplete and Heterogeneous Genomic Data ​

Author: Muhammet Sami Yavuz, Ayhan Can Erdur, Sabri Mustafa Kahya, Benedikt Wiestler, Jana Lipkova
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.GN

arXiv:2607.07725v1 Announce Type: cross Abstract: Genomic prediction models often fail to transfer across institutions because sequencing panels differ across sites, creating structural feature missingness at deployment. Existing approaches to this challenge typically restrict analysis to genes shar...

📖 Read original article


62. Collective Intelligence with Foundation Models ​

Author: J. de Curt`o, I. de Zarz`a
Published: 7/11/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.CL

arXiv:2607.07729v1 Announce Type: cross Abstract: As foundation models grow in scale and diversity, coordinating multiple models into cooperative reasoning systems offers a path toward safer, more reliable AI. This chapter presents a multi-agent framework where solver models generate independent dra...

📖 Read original article


63. Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE ​

Author: Haozhan Tang, Zerui Wang, Yuxian Gu, Song Han, Han Cai
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.07740v1 Announce Type: cross Abstract: Modern LLMs are increasingly deployed in long-context applications such as retrieval-augmented generation, repository-level coding, and agentic workflows whose accumulated reasoning and tool traces routinely push the input an order of magnitude past ...

📖 Read original article


64. Architecture Generalization with MetaNCA ​

Author: Meet Barot, Daniel Berenberg, Sina Khajehabdollahi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.07743v1 Announce Type: cross Abstract: Self-organization is an emergent property of life, driven by the collective behavior of individual components acting on local information. Biological neurons, through local interactions transmitted through synapses, are able to learn efficiently and ...

📖 Read original article


65. A Transdiagnostic Space of Disorder Like Phenotypes in Reinforcement Learning Agents ​

Author: Hari Prasad
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.07753v1 Announce Type: cross Abstract: Modelling psychological disorders in artificial agents offers both a testbed for computational psychiatry and a lens on the failure modes of affective control. Prior work induces one or two disorders in a reinforcement learning (RL) agent by hand-tun...

📖 Read original article


66. Principled Analysis of Deep Reinforcement Learning Evaluation and Design Paradigms ​

Author: Ezgi Korkmaz
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.07769v1 Announce Type: cross Abstract: Starting from the utilization of deep neural networks to approximate the state-action value function that led to winning one of the most challenging games, to algorithmic advancements that allowed solving problems without even explicitly stating the ...

📖 Read original article


67. Graph-Regularized Deep Learning for EEG-Based Emotion Recognition with Psychologically-Grounded Label Structure ​

Author: Dongyang Kuang, Zizheng Ma, Yushan Zhang, Xiaocong Zeng
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.07773v1 Announce Type: cross Abstract: EEG-based emotion recognition is critical for mental health monitoring and affective brain-computer interfaces, yet existing deep learning approaches often treat emotion classes as isolated labels, ignoring their psychological interdependencies. We p...

📖 Read original article


68. From Solvers to Research: Large Language Model-Driven Formal Mathematics at the Research Frontier ​

Author: Eric Jiang, Xiao Liang, Yikai Zhang, Yingjia Wan, Mengting Li, Haikang Deng, Alexander K. Taylor, Justin Baker, Rushil Raghavan, Junyi Zhang, Ying Nian Wu, Andrea L. Bertozzi, Kai-Wei Chang, Raghu Meka, Matthew Sottile, Nanyun Peng, Amit Sahai, Terence Tao, Wei Wang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.07779v1 Announce Type: cross Abstract: Recent developments in AI for Mathematics (AI4Math), especially Large Language Model (LLM)-driven theorem provers, has achieved remarkable success in formal proof generation for well-defined mathematical problems through Interactive Theorem Proving (...

📖 Read original article


69. DreamCharacter-1: From 3D Generative Foundation Models to Product-Ready Character Generation ​

Author: Weizhe Liu, Yunjie Wu, Xiangqian Shu, Guangwei Wang, Xiangyu Xu, Peng Li, Yujie Li, Hengkai Guo
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.07817v1 Announce Type: cross Abstract: We present DreamCharacter-1, a lightweight post-adaptation framework that calibrates pretrained 3D foundation models toward high-fidelity, production-ready 3D character generation. Building upon a 3D foundation backbone, our pipeline incorporates thr...

📖 Read original article


70. From Triggers to Emotions: A CPM-Grounded Appraisal Multi-Agent for Dynamic Emotional Evolution in Persona-Based Dialogue ​

Author: Jingyao Cai, Shuaijun Liu, Abdul Rehman, Yutong Guo, Qin Tian, Thomas Dolby, Sue Green, Chantel Cox, Xiaosong Yang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.MA, cs.AI

arXiv:2607.07824v1 Announce Type: cross Abstract: Large Language Models (LLMs) have substantially advanced persona-based dialogue agents for emotion-sensitive role simulation in healthcare, education, counseling, customer service, and interactive storytelling. However, two related lines of work leav...

📖 Read original article


71. Shift & Drift: A Zero-Shot Benchmark for Generalizable and Robust Autonomous Driving Motion Planning ​

Author: Alessandro Canevaro, Hang Yu, Julian Schmidt, Peizheng Li, Silvan Lindner, Wilhelm Stork, Georg Martius, Julian Jordan
Published: 7/11/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.07844v1 Announce Type: cross Abstract: While closed-loop motion planners trained on large-scale, object-level datasets, e.g., nuPlan, demonstrate strong in-distribution (ID) performance, their generalization to novel urban topologies and recovery mechanisms following execution perturbatio...

📖 Read original article


72. Kime-Representation Formulations of Three Open Problems in the Foundations of Classical Mechanics: Uncertainty, Invariant Entropy, and Directional Degrees of Freedom ​

Author: Ivo D. Dinov
Published: 7/11/2026, 4:00:00 AM
Categories: math-ph, cs.AI, math.MP, physics.comp-ph

arXiv:2607.07851v1 Announce Type: cross Abstract: We give mathematically self-contained formulations, in the complex-time (kime) representation, of three open problems from the foundations of classical mechanics: (I) the extension of the classical entropic uncertainty principle to non-canonical vari...

📖 Read original article


73. Multi-agent Autoformalization of Tensor Network Theory ​

Author: Sirui Lu, Erickson Tjoa, J. Ignacio Cirac
Published: 7/11/2026, 4:00:00 AM
Categories: quant-ph, cs.AI

arXiv:2607.07857v1 Announce Type: cross Abstract: We build a team of specialized large language-model agents and present an agent-driven workflow for research-level formalization in theoretical physics, with the autoformalization of the fundamental theorem of matrix-product states as a demonstration...

📖 Read original article


74. Time-to-Collision Based Dynamic Obstacle Avoidance Using Pretrained Vision Models for Robots in Unstructured Environments ​

Author: Erik Jagnandan, Mulugeta Haile, Gregory Barber, Pratik Chaudhari
Published: 7/11/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, eess.IV

arXiv:2607.07885v1 Announce Type: cross Abstract: Dynamic obstacle avoidance in unstructured outdoor environments remains a critical challenge for autonomous mobile robots, particularly when large-scale robot-specific training data and simulation-based policies are impractical. We present a data-eff...

📖 Read original article


75. How Do I Know What to Say Next? Barenholtz's Autogenerative Theory as an Enrichment of Harrisean Integrationism ​

Author: J. Mark Bishop, Stephen J. Cowley
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.07891v1 Announce Type: cross Abstract: Roy Harris's Integrationist linguistics offers a compelling critique of the referentialist tradition embedded deep at the heart of computational approaches to language, arguing that language is not a code that maps onto a pre-given world but a situat...

📖 Read original article


76. Closed-Loop Dynamic Validator Node Scaling in Private Substrate Blockchains Using Takagi-Sugeno Fuzzy Inference ​

Author: Thandile Nododile, Ayinde M. Usman, Clement N. Nyirenda
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI

arXiv:2607.07901v1 Announce Type: cross Abstract: Private blockchain networks run with fixed node configurations that cannot adapt to changing workload conditions. Too many nodes serving a light workload waste resources; too few nodes facing heavy demand slow block production and degrade finalisatio...

📖 Read original article


77. Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs ​

Author: Anupam Wagle, Ifrat Ikhtear Uddin, Chaowei Zhang, Longwei Wang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI

arXiv:2607.07903v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit remarkable capabilities but remain highly vulnerable to adversarial prompts and jailbreak attacks. Existing approaches primarily analyze these failures through input-output behaviors or attribution methods, offeri...

📖 Read original article


78. Multimodal Unlearning Across Vision, Language, Video, and Audio: Survey of Methods, Datasets, and Benchmarks ​

Author: Nobin Sarwar, Shubhashis Roy Dipta, Zheyuan Liu, Vaidehi Patil
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CR, cs.MM

arXiv:2607.07907v1 Announce Type: cross Abstract: With the growing adoption of VLMs, DMs, LLMs, and AFMs, these multimodal foundation models can inadvertently encode sensitive, copyrighted, biased, or unsafe cross-modal associations that originate from their training data. Retraining after deletion ...

📖 Read original article


79. Efficient Safety Alignment of Language Models via Latent Personality Traits ​

Author: Mohamed Amine Merzouk, Nolan Smyth, Damiano Fornasiere, Linh Le, David Williams-King, Adam Oberman
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CR

arXiv:2607.07918v1 Announce Type: cross Abstract: Current safety methods for large language models are known to be vulnerable to adversarial attacks, motivating research into robust alternatives. Latent Adversarial Training (LAT) is among the most effective defenses, but can degrade utility and requ...

📖 Read original article


80. Adversarial Decoys: Misdirecting Attention-Based Defenses in ViT ​

Author: Giulia Marchiori Pietrosanti, Giulio Rossolini, Giorgio Buttazzo
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.07922v1 Announce Type: cross Abstract: Vision Transformers (ViTs) remain vulnerable to localized adversarial attacks, e.g., adversarial patches, while recent test-time defenses mitigate them by suppressing image tokens with abnormally high attention scores. These defenses exploit a strong...

📖 Read original article


81. path_boost: A Python Package for Interpretable Graph-Level Prediction using Path-Based Gradient Boosting ​

Author: Claudio Meggio, Johan Pensar, Riccardo De Bin
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.07935v1 Announce Type: cross Abstract: We present path_boost, a Python package for interpretable supervised learning on graph-structured input data. The package implements PathBoost, a gradient boosting algorithm that automatically discovers predictive labeled paths within graphs during t...

📖 Read original article


82. Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing ​

Author: Tommaso Cerruti, Tim Rieder, George Rowlands, Lingfeng Jin, Imanol Schlag
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.07953v1 Announce Type: cross Abstract: Self-attention lets each token retrieve information from the full context, but its quadratic cost in sequence length limits training and inference at long context. This paper presents a comparative study of softmax attention and four recent recurrent...

📖 Read original article


83. Beyond Thermal Imaging: Inferring Thermophysical Properties from Time-Resolved Thermal Observations ​

Author: Chenghao Xu, Malcolm Mielle, Olga Fink
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.07962v1 Announce Type: cross Abstract: Inferring latent physical properties from sensory observations is a fundamental challenge in machine perception. Among available sensing modalities, thermal imaging is particularly promising because temperature evolution is directly governed by heat-...

📖 Read original article


84. A Multi-cluster Boundary Learning Method for Out-of-Scope Intent Detection via MiniLM Embedding ​

Author: Yihong Xu, Mingyu Kang, Linyuan L"u
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.07974v1 Announce Type: cross Abstract: Intent detection is a critical task that bridges human intents and system actions in human-machine interaction systems. However, there still exist challenges for detecting out-of-scope (OOS) intents. (i) The traditional methods view the OOS intent de...

📖 Read original article


85. When Implausible Tokens Get Reinforced: Tail-Aware Credit Calibration for LLM Reinforcement Learning ​

Author: Xiuyi Lou, Zicheng Xu, Yu-Neng Chuang, Hoang Anh Duy Le, Zhaozhuo Xu, Guanchu Wang, Vladimir Braverman
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.07976v1 Announce Type: cross Abstract: Reinforcement learning (RL) has achieved remarkable success in enhancing the reasoning capabilities of large language models (LLMs). However, widely used critic-free RL methods rely on uniform credit assignment, broadcasting the same advantage to all...

📖 Read original article


86. 3100 Opinions on Code Review in an AI World: Building Causal Theory from Practitioner Discourse ​

Author: Shyam Agarwal, Courtney Miller, Christian K"astner, Bogdan Vasilescu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.SE, cs.AI

arXiv:2607.07980v1 Announce Type: cross Abstract: Coding agents now author entire pull requests, and practitioners sharply disagree about what this does to code review: whether it becomes the bottleneck, whether human review is still necessary, and whether it quietly erodes the understanding that it...

📖 Read original article


87. A Reliability Assessment of LALM Audio Judges for Full-Duplex Voice Agents ​

Author: A. Sayyad, J. Emmons, S. Jones, T. Lin, H. Krishnan
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.SD, eess.AS

arXiv:2607.07985v1 Announce Type: cross Abstract: We report the empirical reliability of Gemini models as audio judges that score full-duplex agent conversations directly from the raw stereo waveform, tested across three models in the Gemini family: 2.5 Flash, 3.5 Flash, and 3.1 Pro. Our primary evi...

📖 Read original article


88. Who Broke the System? Failure Localization in LLM-Based Multi-Agent Systems ​

Author: Yufei Xia, Anjun Gao, Yueyang Quan, Zhuqing Liu, Minghong Fang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.IR, cs.LG, cs.MA

arXiv:2607.07989v1 Announce Type: cross Abstract: Large language model (LLM) based multi-agent systems enable complex problem solving through coordinated reasoning and action, but their distributed structure also introduces new challenges in diagnosing system-level failures. When an execution fails,...

📖 Read original article


89. SpO$_2$ Predictor-Guided Stage-Wise Time-Frequency Reconstruction of Low-Quality Dual-Wavelength PPG for Oxygen Saturation Estimation ​

Author: Zequan Liang, Elahe Hosseini, Ning Miao, Mahdi Pirayesh Shirazi Nejad, Wei Shao, Ehsan Kourkchi, Setareh Rafatirad, Houman Homayoun
Published: 7/11/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG

arXiv:2607.07996v1 Announce Type: cross Abstract: Continuous oxygen saturation (SpO$_2$) estimation from wearable photoplethysmography (PPG) is important for long-term health monitoring, but low-quality red and infrared PPG segments can distort waveform morphology and degrade SpO$_2$ prediction accu...

📖 Read original article


90. Reaction-network reasoning with frontier models for experimentally confirmed catalyst-selectivity hypotheses ​

Author: Sutanay Choudhury, Anwesha Banerjee, Udishnu Sanyal, Jorin Dawidowicz, Chiezugolum Ijeoma Odilinye, Jesun Firoz, Liney Arnadottir, Simone Raugei, Johannes Lercher, Arnab Dutta
Published: 7/11/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.AI

arXiv:2607.08003v1 Announce Type: cross Abstract: Catalysts are essential for sustainable chemical manufacturing, yet discovering novel architectures remains a bottleneck dominated by trial-and-error experimentation and computationally intensive screening. In complex reactions such as electrochemica...

📖 Read original article


91. Beware What You Autocomplete: Forensic Attribution of Backdoored Code Completions ​

Author: Anjun Gao, Yueyang Quan, Zhuqing Liu, Minghong Fang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.IR, cs.LG

arXiv:2607.08011v1 Announce Type: cross Abstract: Large language models have enabled powerful code completion systems that assist developers by predicting subsequent lines of code. However, these models remain vulnerable to backdoor attacks, where malicious fine-tuning data covertly implants unsafe ...

📖 Read original article


92. Provably Optimal Learning Algorithms for Assistance Games ​

Author: Nivasini Ananthakrishnan, Mark Bedaywi, Michael I. Jordan, Stuart Russell, Nika Haghtalab
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GT

arXiv:2607.08012v1 Announce Type: cross Abstract: This paper studies an online variant of the assistance games framework, where an informed agent and an uninformed agent repeatedly interact over $T$ timesteps to optimize a common reward function. While the informed agent (the human) observes a laten...

📖 Read original article


93. Can We Trust LLM's Logic? Quantifying Uncertainty, Coherence, and Robustness via a Graph-Based Framework ​

Author: Riccardo Revalor, Jalees Rehman, Debjit Pal
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.08017v1 Announce Type: cross Abstract: Large-Language Models (LLMs) can be prone to flawed and unfaithful reasoning that decoding strategies like Self-Consistency (SC) fail to detect as they evaluate only final-answer agreement while ignoring the logical validity of intermediate steps. Th...

📖 Read original article


94. APIVOT: Adaptive Planning with Interleaved Vision-Language Thoughts ​

Author: Emily Jin, Joy Hsu, Yiqing Xu, Weiyu Liu, Nick Haber, Jiajun Wu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.RO

arXiv:2607.08024v1 Announce Type: cross Abstract: Long-horizon robot planning requires jointly reasoning over semantic task structure and geometric feasibility. To successfully execute a task, a robot must decompose goals, select task-relevant objects, and sequence actions, while ensuring that plans...

📖 Read original article


95. Structured Pruning of Large Language Models via Power Transformation and Sign-Preserving Score Aggregation with Adaptive Feature Retention ​

Author: Ryota Kobayashi, Tsubasa Hirakawa, Takayoshi Yamashita, Hironobu Fujiyoshi, Yasunori Ishii, Tomoyuki Okuno, Kazuki Kozuka
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.08027v1 Announce Type: cross Abstract: This paper proposes an improved structured pruning method for large language models (LLMs) that addresses key challenges in adapting Adaptive Feature Retention (AFR), an unstructured pruning technique, to structured pruning. When applying AFR to stru...

📖 Read original article


96. DKDNet: Dual Knowledge and Data-Driven Network for Cross-Domain Automatic Modulation Classification ​

Author: Shuang Wang, Chenxu Wang, Hantong Xing, Hanlin Mo, Lirong Han, Licheng Jiao
Published: 7/11/2026, 4:00:00 AM
Categories: eess.SP, cs.AI

arXiv:2607.08031v1 Announce Type: cross Abstract: The dynamics of communication environments induce significant distribution shifts across domains, challenging the generalization of deep learning-based automatic modulation classification (AMC) models. While existing UDA methods alleviate this proble...

📖 Read original article


97. PLURAL: A Global Dataset for Value Alignment ​

Author: Dhruv Agarwal, Anya Shukla, Tanya Goyal, Aditya Vashistha
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY

arXiv:2607.08034v1 Announce Type: cross Abstract: Large language models (LLMs) are used worldwide, yet disproportionately reflect Western values, limiting their ability to represent diverse value systems. We introduce PLURAL, a large-scale, value-focused preference dataset grounded in the Integrated...

📖 Read original article


98. Aleena: Alignment Agent for Research Software Engineering Collaborations ​

Author: Kshitij Dani, Cordero Core, Landung Setiawan, Carlos Garcia Jurado Suarez, Anshul Tambay, Vani Mandava, Anant Mittal
Published: 7/11/2026, 4:00:00 AM
Categories: cs.SE, cs.AI

arXiv:2607.08043v1 Announce Type: cross Abstract: Research software collaborations span meetings, informal chats, pull requests, and GitHub issues. A decision surfaced in a Slack thread, refined in a meeting, and implemented in a pull request can lose its original rationale across these artifacts, l...

📖 Read original article


99. What LLM Forecasters Know but Don't Say: Probing Internal Representations for Calibration and Faithfulness ​

Author: Rapha"el Sarfati, Pratyush Ranjan Tiwari, Siddharth Boppana, Christopher J. Earls, Srikar Varadaraj, Eric Ho
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.08046v1 Announce Type: cross Abstract: Large language models fine-tuned for forecasting can be accurate yet poorly calibrated, and their chain-of-thought (CoT) reasoning may not faithfully reflect the evidence behind a forecast. We ask whether internal representations offer a more direct ...

📖 Read original article


100. Who Analyses the Analyser? Self-Validating LLM Hazard Analysis with Constitutional Meta-STPA ​

Author: Samuel Tetteh, Udip Shrestha, Joshua R. Waite, Cody Fleming
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08054v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly trusted to draft the artifacts of safety analysis such as, losses, hazards, Unsafe Control Actions (UCAs), and safety constraints, inside rigorous processes such as Systems-Theoretic Process Analysis (STP...

📖 Read original article


101. Reinforcing the Generation Order of Multimodal Masked Diffusion Models ​

Author: Yidong Ouyang, Zhe Wang, Sourav Bhabesh, Dmitriy Bespalov
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.08056v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) have recently achieved substantial progress in natural language generation tasks. Recent research demonstrates that adaptive token generation ordering can significantly improve performance in mathematical reasoning an...

📖 Read original article


102. Towards Efficient Large Language Model Serving: A Survey on System-Aware KV Cache Optimization ​

Author: Jiantong Jiang, Peiyu Yang, Rui Zhang, Feng Liu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.08057v1 Announce Type: cross Abstract: Despite the rapid advancements of large language models (LLMs), LLM serving systems remain memory-intensive and costly. The key-value (KV) cache, which stores KV tensors during autoregressive decoding, is crucial for enabling low-latency, high-throug...

📖 Read original article


103. When Thinking Hurts: Epistemic Signals in the Reasoning Chains of Visual Language Models ​

Author: Mayank Singal
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08059v1 Announce Type: cross Abstract: Uncertainty quantification for visual language models (VLMs) conventionally targets the answer token distribution. We provide the first three-family empirical characterisation of answer entropy behaviour in thinking-mode VLMs. Running four models on ...

📖 Read original article


104. COBART: Controlled, Optimized, Bidirectional and Auto-Regressive Transformer for Ad Headline Generation ​

Author: Yashal Shakti Kanungo, Gyanendra Das, Pooja A, Sumit Negi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.08071v1 Announce Type: cross Abstract: Online ads are essential to all businesses and ad headlines are one of their core creative component. Existing methods can generate headlines automatically and also optimize their click-through-rate (CTR) and quality. However, evolving ad formats and...

📖 Read original article


105. LDFE: Laplacian Decoupled Feature Enhancement Block for Dual-Stream CNN-based RGB-IR Object Detection ​

Author: Wenhao Dong, Xiaoyan Luo, Linlin Yang, Haodong Zhu, Xiaorong Shi, Guodong Guo, Baochang Zhang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.08076v1 Announce Type: cross Abstract: The complementary information between RGB and IR images can significantly enhance object detection performance under extreme conditions. Existing methods prefer dual-stream CNN backbones built upon YOLO for feature extraction and focus on the design ...

📖 Read original article


106. Deep Learning Method for Stationary Distribution of Reflected Brownian Motion ​

Author: Jim Dai, Zhanhao Zhang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08091v1 Announce Type: cross Abstract: The stationary distribution of reflected Brownian motion (RBM) plays an important role in the analysis of high-dimensional stochastic systems, yet closed-form solutions are known only for a few special cases. Computing important performance metrics, ...

📖 Read original article


107. PS4: Proxy-Supervised Joint Training for Real Target Speaker Extraction ​

Author: Wanyi Ning, Wei Zhou, Yingpeng Li, Yinshang Guo, Haitao Qian, Yiming Cheng
Published: 7/11/2026, 4:00:00 AM
Categories: cs.SD, cs.AI

arXiv:2607.08111v1 Announce Type: cross Abstract: Training target speaker extraction (TSE) models for real conversational mixtures remains challenging because large-scale training corpora and clean target speech for supervision are unavailable. We present PS4, a proxy-supervised training framework f...

📖 Read original article


108. ICDAR 2026 HIPE-OCRepair Competition on LLM-Assisted OCR Post-Correction for Historical Documents ​

Author: Maud Ehrmann, Emanuela Boros, Juri Opitz, Andrianos Michail, Florian Wagner, Simon Clematide
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.IR

arXiv:2607.08143v1 Announce Type: cross Abstract: We present the results of HIPE-OCRepair-2026, an ICDAR competition on LLM-assisted OCR post-correction of historical documents. OCR post-correction remains a long-standing challenge in digital heritage: large-scale collections of digitized documents ...

📖 Read original article


109. Prismata: Confining Cross-Site Prompt Injection in Web Agents ​

Author: Corban Villa, Alp Eren Ozdarendeli, Sijun Tan, Raluca Ada Popa
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI

arXiv:2607.08147v1 Announce Type: cross Abstract: Autonomous web agents promise to automate everyday browsing tasks, but inherit one of the web's oldest attack surfaces. Cross-Site Scripting proved that mixing trusted and untrusted content is dangerous, even on benign pages. Agents resurface this ri...

📖 Read original article


110. LEXIC: Lightweight Eye-tracking eXtension via Injected Complexity ​

Author: Sumin Lee, Kyeonghun Kim, Subeen Lee, Jiwon Yang, Tien Nguyen, Ken Ying-Kai Liao, Nam-Joon Kim
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC, cs.LG

arXiv:2607.08152v1 Announce Type: cross Abstract: On the recent EyeBench benchmark, predicting reading comprehension from eye movements exposes a stark gap: text-aware models using pretrained language models reach 56--63% AUROC, while gaze-only models operate at chance. We ask how far a gaze-only mo...

📖 Read original article


111. ProsMAE: Multi-Source MAE Pretraining for ISUP Grade Classification ​

Author: Anna Jung, Kyeonghun Kim, Youngung Han, Eunseob Choi, Jiwon Yang, Ken Ying-Kai Liao, Hyuk-Jae Lee, Nam-Joon Kim
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.08162v1 Announce Type: cross Abstract: Whole slide images (WSIs) provide rich diagnostic information for computational pathology, but their gigapixel scale, stain variation, scanner differences, tissue artifacts, and limited expert annotation make robust model training challenging. This p...

📖 Read original article


112. Out of Sight: Compression-Aware Content Protection against Agentic Crawlers ​

Author: Xuefei Wang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI

arXiv:2607.08180v1 Announce Type: cross Abstract: The rise of LLM-based agents with reasoning, summarization, and memory capabilities has created a new threat surface for online content that conventional defenses fail to address. Existing defenses like access controls can be circumvented by agents m...

📖 Read original article


113. LEEVLA: Seeing What Matters in Latent Environment Evolution for Vision-Language-Action ​

Author: Qi Lyu, Baicheng Liu, Xudong Wang, Jiahua Dong, Lianqing Liu, Zhi Han
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.08182v1 Announce Type: cross Abstract: Vision-language-action (VLA) models aim to map multimodal inputs to robot actions. However, most existing approaches struggle to cover complex dynamic scenarios due to treating all visual tokens uniformly and reasoning with human-selected factors, wh...

📖 Read original article


114. Leveraging Color Naming for Image Enhancement ​

Author: David Serrano-Lozano, Luis Herranz, Michael S. Brown, Javier Vazquez-Corral
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.08185v1 Announce Type: cross Abstract: Enhancing images to make them visually appealing is a persistent challenge in computer vision. Many deep-learning methods train models on paired datasets to replicate expert editing styles. However, these approaches struggle with two key issues: (1) ...

📖 Read original article


115. Open-ended Multi-agent Autocurricula via Visual Inspection of Policies with Multi-modal LLMs ​

Author: Lorenzo Pant`e, Andrea Fanti, Roberto Capobianco
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08193v1 Announce Type: cross Abstract: Open-ended curricula in Reinforcement Learning (RL) aim to train generally-capable agents by identifying tasks that facilitate learning increasingly complex skills. A major challenge when designing such curricula is assessing task difficulty relative...

📖 Read original article


116. TMI: Text-to-Image Meets Image-to-Image for Complementary Data Synthesis to Boost Long-Tailed Instance Segmentation ​

Author: Hyeonseop Song, Seokhun Choi, Hoseok Do
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.08201v1 Announce Type: cross Abstract: Large-vocabulary instance segmentation is constrained by long-tailed category distributions and fine-grained inter-class ambiguity. While data synthesis offers a promising alternative, current paradigms have complementary limitations: text-to-image (...

📖 Read original article


117. RhyMix: A Lightweight Adaptive Multi-Rhythm Network for Long-Term Time Series Forecasting ​

Author: Sumit Satishrao Shevtekar, Chandresh Kumar Maurya
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08234v1 Announce Type: cross Abstract: Real-world time series exhibit complex dynamics characterized by multiple simultaneous temporal patterns: short-term fluctuations, periodic seasonal cycles, long-term trends, and irregular abrupt changes. However, many existing forecasting architectu...

📖 Read original article


118. Best-of-$N$ TTS Evaluation is Confounded by ASR Family Alignment ​

Author: Taehyung Yu, Seongjae Kang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.SD

arXiv:2607.08256v1 Announce Type: cross Abstract: Best-of-$N$ (BoN) inference improves content consistency in zero-shot text-to-speech by selecting from $N$ candidates with an automatic speech recognition (ASR) verifier. We identify an underexplored evaluation confound: a verifier's apparent quality...

📖 Read original article


119. Multi-Agent Firewall Architecture for Privacy Protection of Sensitive Data in Interactions with Language Models ​

Author: Hugo Garc'ia Cuesta, Pablo Mateo Torrej'on, Alfonso S'anchez-Maci'an
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.MA

arXiv:2607.08282v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have become essential productivity tools, their integration into workflows without adequate safeguards creates significant risks. This paper proposes an open-source, privacy-focused, user-facing firewall designed to...

📖 Read original article


120. From Legacy Documentation to OSCAL: An MCP-Based Agent Pipeline for Threat-Informed Continuous Compliance in Critical Infrastructure ​

Author: Lea Roxanne Muth, Marian Margraf
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI

arXiv:2607.08288v1 Announce Type: cross Abstract: In critical infrastructure, operational technology environments often cannot be actively scanned, and yet active system feedback is needed for risk assessment and compliance. This paper presents a non-invasive, MCP-grounded multi-agent pipeline that ...

📖 Read original article


121. GitLake: Git-for-data for the agentic lakehouse ​

Author: Weiming Sheng, Jinlang Wang, Manuel Barros, Aldrin Montana, Jacopo Tagliabue, Luca Bigon
Published: 7/11/2026, 4:00:00 AM
Categories: cs.DB, cs.AI

arXiv:2607.08319v1 Announce Type: cross Abstract: We present GitLake, a Git-for-data design for an agent-first lakehouse. The system lifts single-table Iceberg snapshots into lakehouse-wide commits, branches, and merges, letting agents work on isolated branches while humans review and publish change...

📖 Read original article


122. ArtMine: Discovering and Formalizing Artistic Processes ​

Author: Kaustubh Kumar, Ashutosh Ranjan, Vivek Srivastava, Blessin Varkey, Shirish Karande
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08331v1 Announce Type: cross Abstract: Understanding how artworks are created requires reasoning about the iterative decisions, material operations, and contextual influences that shape artistic production. While recent generative AI systems can synthesize artworks with high fidelity, the...

📖 Read original article


123. TypeProbe: Recovering Type Representations from Hidden States of Pre-trained Code Models ​

Author: Giuliano Gorgone, Fausto Carcassi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.PL

arXiv:2607.08339v1 Announce Type: cross Abstract: State-of-the-art code models achieve impressive performance, yet the extent to which they internally encode type information remains poorly understood. We probe the residual streams of pretrained code models for internal type representations using a ...

📖 Read original article


124. Spectral Analysis of Dueling Q-Learning ​

Author: Donghwan Lee
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08340v1 Announce Type: cross Abstract: Q-learning is a fundamental algorithm in reinforcement learning (RL) for solving discounted Markov decision processes (MDPs) when the transition kernel is unknown. The deep Q-network (DQN) extends Q-learning by using a deep neural network for Q-funct...

📖 Read original article


125. FSD-VLN: Fast-Slow Dual-System Modeling for Aerial Long-Horizon Vision-Language Navigation ​

Author: Xueke Zhu, Qingyan Meng, Liutao Yu, Wei Zhang, Zhengyu Ma, Huihui Zhou, Yonghong Tian
Published: 7/11/2026, 4:00:00 AM
Categories: cs.RO, cs.AI

arXiv:2607.08359v1 Announce Type: cross Abstract: Vision-Language Navigation (VLN) enables UAV autonomous navigation in unknown environments by mapping language instructions to real-time visual inputs. Compared with GPS-dependent or pre-programmed navigation, VLN supports intuitive human-machine int...

📖 Read original article


126. On the Role of Conversational Timing in Synthetic Training Data for ASR ​

Author: M'at'e Gedeon, P'eter Mihajlik
Published: 7/11/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.SD

arXiv:2607.08371v1 Announce Type: cross Abstract: Synthetic multi-speaker conversations are widely used to train conversational automatic speech recognition (ASR) systems, but it remains unclear which timing properties make simulated data most useful. This paper studies conversational timing as a co...

📖 Read original article


127. Self-Adaptive Anomaly Detection with Reinforcement Learning and Human Feedback in Connected Vehicles ​

Author: Matthias Wei{\ss}, Athreya Hosahalli Prakash, Maurice Artelt, Falk Dettinger, Nasser Jazdi, Michael Weyrich
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08373v1 Announce Type: cross Abstract: Connected vehicles are autonomous cyber-physical systems whose behavior must be continuously monitored during operation to detect deviations from normal operation before they propagate into failures. Such evaluation is challenging because the systems...

📖 Read original article


128. Large-Language-Models-as-a-Judge in Theory-Agnostic Adaptive Metric-Alignment for Prototypical Networks in Personality Recognition ​

Author: Jing Jie Tan, Ban-Hoe Kwan, Danny Wee-Kiat Ng, Yan-Chai Hum, Shih-Yu Lo, Po-An Chen, Noriyuki Kawarazaki, Kosuke Takano, Anissa Mokraoui
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC, cs.RO, cs.SI

arXiv:2607.08374v1 Announce Type: cross Abstract: Personality recognition has traditionally been constrained by theory-dependent formulations, where models are trained to fit predefined psychological taxonomies rather than uncovering shared underlying behavioral structure. This limits generalization...

📖 Read original article


129. WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving ​

Author: Xuerun Yan, Zhexi Lian, Nuoheng Zhang, Shiyu Fang, Haoran Wang, Chen Lv, Jia Hu, Binyang Song
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.08375v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving. However, existing methods either lack comprehensive world cognition or suffer from fragmented world foresight, inherently confining these models to reactive driving. To ...

📖 Read original article


130. TRACE: A Two-Channel Robust Attribution Watermark via Complementary Embeddings for LLM-Agent Trajectories ​

Author: Zheng Gao, Xiaoyu Li, Xiaoyan Feng, Jiaojiao Jiang, Yang Song, Yulei Sui, Zhenchang Xing, Liming Zhu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2607.08400v1 Announce Type: cross Abstract: LLM agents reach users through resellers, who may rebrand a developer's agent or substitute a cheaper model. When provenance is disputed, attribution rests on the trajectory log (the record of tool calls, observations, and executed actions, not the m...

📖 Read original article


131. Swapping Faces, Saving Features: A Dual-Purpose Pipeline for Pedestrian Privacy in ITS ​

Author: Roba H. Farouk, Catherine M. Elias
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.RO

arXiv:2607.08402v1 Announce Type: cross Abstract: Large-scale and diverse datasets are needed to train AI models to take real-time decisions for autonomous vehicles (AVs), an intelligent transportation system (ITS) application. Pedestrian intention and trajectory prediction are critical models used ...

📖 Read original article


132. DrugGen 2: A disease-aware language model for enhancing drug discovery ​

Author: Ali Motahharynia, Mohammadreza Ghaffarzadeh-Esfahani, Mahsa Sheikholeslami, Navid Mazrouei, Matin Irajpour, Yousof Gheisari, Hajar Sirous
Published: 7/11/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG

arXiv:2607.08404v1 Announce Type: cross Abstract: Current computational approaches for drug design typically focus on generating molecules conditioned on specific targets or general molecular properties, often neglecting the influence of disease context on target behavior and therapeutic outcomes. T...

📖 Read original article


133. Track2Map: Online Deformable SLAM with Motion-Aware Pose Optimization in Robotic Surgery ​

Author: Tianyi Song, Sierra Bonilla, Xinwei Ju, Evangelos Mazomenos, Danail Stoyanov, Adam Schmidt, Omid Mohareri, Sophia Bano, Francisco Vasconcelos
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.08408v1 Announce Type: cross Abstract: Gaussian splatting is the current state-of-the-art for dense, deformable 3D anatomy reconstruction in robot-assisted minimally invasive surgery (RAMIS); however, most pipelines are offline and depend on accurate camera trajectory priors (often from r...

📖 Read original article


134. When Synthetic Speech Is All You Have: Better Call GRPO ​

Author: Shashi Kumar, Yanis Labrak, Hasindri Watawana, Sergio Burdisso, Esa'u Villatoro-Tello, Kadri Hacio\u{g}lu, Petr Motlicek, Andreas Stolcke
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.08409v1 Announce Type: cross Abstract: LLM-based ASR adapted to regulated domains such as banking is bottlenecked by privacy: real speech is costly and legally constrained to collect, making synthetic text-to-speech (TTS) an attractive substitute. Yet synthetic speech stays acoustically m...

📖 Read original article


135. Predicting Male Fertility Using Machine Learning: A Semen Parameters Based Analysis with the VISEM Dataset ​

Author: Shahnawaz Qureshi, Raja Khurram Shahzad, Muhammad Fozan, Emal Kawal, Syed Aziz Shah, Sattam Al-Anazi, Syed MuhammadZeeshan Iqbal
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08429v1 Announce Type: cross Abstract: Male infertility is a significant yet often underdiagnosed aspect of reproductive health, with semen analysis serving as the cornerstone of clinical evaluation. To address this problem, this study investigates the use of machine learning algorithms t...

📖 Read original article


136. EgoWAM: World Action Models Beyond Pixels with In-the-Wild Egocentric Human Data ​

Author: Baoyu Li, Xinchen Yin, Mengying Lin, Yixin Zhang, Danfei Xu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.RO, cs.AI

arXiv:2607.08436v1 Announce Type: cross Abstract: Egocentric human data offers scalable supervision for robot manipulation. However, behavior cloning entangles transferable content like objects, scenes, and task semantics, with non-transferable factors like human morphology, head motion, and behavio...

📖 Read original article


137. ADORN: Adaptive Drift handling for Open RAN using Reinforcement Learning ​

Author: Ashit Kumar Subudhi, Bhargav Chirumamilla, Shubham Vaishnav, Mduduzi C. Hlophe, Praveen Kumar Donta, Andrea Fumagalli, Venkateswarlu Gudepu, Koteswararao Kondepu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.NI, cs.AI

arXiv:2607.08443v1 Announce Type: cross Abstract: Dynamic traffic variations in Open Radio Access Networks (O-RAN) lead to drift, which degrades the performance of Artificial Intelligence/Machine Learning (AI/ML) models. Traditional retraining approaches maintain forecasting accuracy but incur high ...

📖 Read original article


138. Spatio-Temporal Scheduling Prediction Under Backhaul Delay for Resilient Coordinated Beamforming ​

Author: Prashant Kumar Singh, Shubham Vaishnav, Ahmet Hasim G"okceoglu, Li Wang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.NI, cs.AI

arXiv:2607.08454v1 Announce Type: cross Abstract: Coordinated beamforming in distributed 5G networks relies on the timely exchange of inter-cell scheduling information, but backhaul latency makes this information stale. Even a single transmission time interval (TTI) of delay can reduce CBF-SLNR perf...

📖 Read original article


139. Two Axes of LLM Abstention: Answer Correctness and Question Answerability ​

Author: Benedikt J. Wagner
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.08456v1 Announce Type: cross Abstract: A model should refuse two different things: answers it would get wrong, and questions it should not answer at all, such as unanswerable ones or ones resting on a false premise. The usual recipe thresholds a single confidence score, which cannot tell ...

📖 Read original article


140. VEGAS: Human-Aligned Video Caption Evaluation via Gaze ​

Author: Shenghui Chen, Po-han Li, Ximeng Sun, Shijia Yang, Emad Barsoum, Zicheng Liu, Sandeep Chinchali, Ufuk Topcu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.HC

arXiv:2607.08489v1 Announce Type: cross Abstract: Vision-language models excel at video captioning, yet typically generate descriptions that fail to capture individual viewers' attention. We propose VEGAS (Video caption Evaluation via GAze Score), a training-free metric that leverages test-time gaze...

📖 Read original article


141. The Context Access Divide: Interaction-Level Architecture as a Complementary Dimension of Agentic Inequality ​

Author: Masahiro Fujita
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CY, cs.AI

arXiv:2607.08495v1 Announce Type: cross Abstract: Sharp et al. (2025) introduce "agentic inequality" as a framework for analyzing disparities in access to AI agents across three dimensions: availability, quality, and quantity. These person- and organization-level dimensions characterize who can acce...

📖 Read original article


142. Cognitive-structured Multimodal Agent for Multimodal Understanding, Generation, and Editing ​

Author: Feng Wang, Canmiao Fu, Zhipeng Huang, Chen Li, Jing Lyu, Ge Li
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG

arXiv:2607.08497v1 Announce Type: cross Abstract: Recent unified multimodal models show a single architecture can jointly perform vision/language understanding and image generation/editing. However, they repeatedly feed all historical visual and textual inputs into a shared context window, limiting ...

📖 Read original article


143. When the Judge Changes, So Does the Measurement: Auditing LLM-as-Judge Reliability ​

Author: Zongyou Yang, Yinghan Hou, Xiaokun Yang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.08535v1 Announce Type: cross Abstract: An LLM-as-judge score can move even when the candidate responses stay fixed, simply because the evaluator has changed. We treat this evaluator-replacement ambiguity as a measurement-validity problem. Across four judgment datasets, we compare two upgr...

📖 Read original article


144. DocMaster: A Hierarchical Structure-Aware System for Document Analysis ​

Author: Ziqi Chen, Yingli Zhou, Fangyuan Zhang, Quanqing Xu, Chuanhui Yang, Yixiang Fang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.DB, cs.AI

arXiv:2607.08539v1 Announce Type: cross Abstract: Leveraging large language models (LLMs) to analyze complex documents -- such as academic papers, technical manuals, and financial reports -- has emerged as a mainstream and critical task in both research and industry. In practice, users must first fi...

📖 Read original article


145. VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval ​

Author: ZhiXin Sun
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.08541v1 Announce Type: cross Abstract: Open-vocabulary object detection and segmentation aim to recognize arbitrary objects beyond predefined categories. Although recent vision-language and reference-based approaches have significantly advanced this field, they often rely on text prompts,...

📖 Read original article


146. SMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric Scheduling ​

Author: Jiahao Wang, Kaizhan Lin, Kaixi Zhang, Jinbo Han, Xingda Wei, Sijie Shen, Chenguang Fang, Wenyuan Yu, Rong Chen, Haibo Chen
Published: 7/11/2026, 4:00:00 AM
Categories: cs.DC, cs.AI

arXiv:2607.08565v1 Announce Type: cross Abstract: LLM scheduling is critical to serving, yet it remains unclear how well existing designs fit agentic serving--with LLM requests issued by agents instead of humans. This shifts the workload in two ways: (1) agents act only on complete responses, making...

📖 Read original article


147. When Structured Sparse Autoencoders Learn Consistent Concepts Across Modalities ​

Author: Weiduo Liao, Yunqiao Yang, Ying Wei
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.08605v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have emerged as a promising technique for mechanistic interpretability by learning a set of sparse latent features in large models, each of which encodes a distinct concept. However, in vision-language models (VLMs), vanill...

📖 Read original article


148. UltraX: Refining Pre-Training Data at Scale with Adaptive Programmatic Editing ​

Author: Xinlong Zhao, Dongsheng Liu, Hengyu Zhao, Zixuan Fu, Zheng Wang, Jie Cai, Jie Zhou, Qiang Ma, Xuanhe Zhou, Xu Han, Yudong Wang, Zhiyuan Liu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.08646v1 Announce Type: cross Abstract: As available training data approaches its physical limit, gains from Scaling Laws have begun to diminish. Consequently, improving Large Language Models (LLMs) now depends less on data expansion and more on higher-quality data utilization. However, in...

📖 Read original article


149. Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning ​

Author: Ali Larian, Qian Lin, Chang Zong Wu, Daniel S. Brown
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08647v1 Announce Type: cross Abstract: As autonomous agents are increasingly deployed across diverse operational contexts, aligning their behavior with human intent demands reward functions that remain robust to such changes rather than overfitting to any single environment. Inverse reinf...

📖 Read original article


Author: Xiaoshuai Song, Liancheng Zhang, Kangzhi Zhao, Yutao Zhu, Zhongyuan Wang, Guanting Dong, Jinghan Yang, Han Li, Kun Gai, Ji-Rong Wen, Zhicheng Dou
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.MA

arXiv:2607.08662v1 Announce Type: cross Abstract: Large language model (LLM)-based web search agents are transforming information seeking from simple factoid question answering into complex, deep-and-wide search and research-oriented tasks. A single ReAct-style agent is constrained by one long traje...

📖 Read original article


151. A Practical Investigation of Training-free Relaxed Speculative Decoding ​

Author: Guoxuan Xia, Luka Ribar, Paul Balanca
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08690v1 Announce Type: cross Abstract: Speculative decoding accelerates sampling from an autoregressive LLM by using a faster auxiliary model to draft tokens which are then verified in parallel by the LLM. Standard speculative decoding is lossless: its rejection and resampling steps exact...

📖 Read original article


152. ProjAgent: Procedural Similarity Retrieval for Repository-Level Code Generation ​

Author: QiHong Chen, Aaron Imani, Iftekhar Ahmed
Published: 7/11/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.IR

arXiv:2607.08691v1 Announce Type: cross Abstract: Repository-level code generation requires implementing target functions while accounting for complex cross-file dependencies and project-specific conventions. Existing retrieval methods predominantly rely on lexical, structural, or semantic similarit...

📖 Read original article


153. Pose-to-Biomechanics: Bridging 3D Human Pose Estimation and Biomechanical Attribute Prediction ​

Author: Ayda Eghbalian, Kevin Desai
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.08725v1 Announce Type: cross Abstract: Recent progress in 3D human pose estimation has made markerless recovery of skeletal motion increasingly accurate and scalable. However, most pose estimators remain optimized for geometric keypoint accuracy, while many real-world applications in reha...

📖 Read original article


154. Validity of LLMs as data annotators: AMALIA on authority ​

Author: Manuel Pita
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY

arXiv:2607.08731v1 Announce Type: cross Abstract: A national language model offers a linguistic community its own instrument for measuring what its citizens say and value. Portugal's AMALIA, a publicly funded 9B-parameter model for European Portuguese, appears competitive on agreement alone: asked t...

📖 Read original article


155. Dimensionality Reduction Meets Network Science: Sensemaking on UMAP's kNN Graph ​

Author: Duen Horng Chau, Donghao Ren, Fred Hohman, Dominik Moritz
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DS, cs.HC

arXiv:2607.08746v1 Announce Type: cross Abstract: While UMAP is widely used for exploring high-dimensional data, typical workflows focus on its lower-dimensional embedding, largely overlooking the rich k-nearest-neighbor (kNN) graph that UMAP constructs internally. This graph encodes the data manifo...

📖 Read original article


156. SLORR: Simple and Efficient In-Training Low-Rank Regularization ​

Author: David Gonz'alez-Mart'inez, Shiwei Liu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.08754v1 Announce Type: cross Abstract: Low-rank factorization is widely used to compress neural networks, but modern models are often not naturally amenable to aggressive factorization without significant accuracy loss. Existing training-time low-rank regularizers can improve compressibil...

📖 Read original article


157. OpenCoF: Learning to Reason Through Video Generation ​

Author: Xinyan Chen, Ziyu Guo, Renrui Zhang, Dongzhi Jiang, Hongsheng Li
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.08763v1 Announce Type: cross Abstract: Reasoning has become a core capability for large models, especially when reliable decisions require understanding logical consequences. Recent video generation models offer a reasoning path distinct from previous Chain-of-Thought (CoT): reasoning can...

📖 Read original article


158. IFAR: Multi-Perspective and Multi-Level Causal Discovery with LLMs ​

Author: Jinwei He, Feng Lu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2409.05559v2 Announce Type: replace Abstract: Large language models (LLMs) have developed rapidly, and their reasoning capabilities have become a hot research topic. However, there is still limited exploration of abductive reasoning. The multi-perspective and multi-level of causes is one of th...

📖 Read original article


159. Goal-Driven Reasoning in DatalogMTL with Magic Sets ​

Author: Shaoyu Wang, Kaiyue Zhao, Dongliang Wei, Przemys{\l}aw Andrzej Wa{\l}\k{e}ga, Dingmin Wang, Hongming Cai, Pan Hu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2412.07259v5 Announce Type: replace Abstract: DatalogMTL is a powerful rule-based language for temporal reasoning. Due to its high expressive power and flexible modeling capabilities, it is suitable for a wide range of applications, including tasks from industrial and financial sectors. Howeve...

📖 Read original article


160. Dual-Difficulty Curriculum Learning for Direct Preference Optimization ​

Author: Mengyang Li, Haozhan Geng, Zhong Zhang, Shuang Liu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2504.07856v4 Announce Type: replace Abstract: Curriculum learning enhances Direct Preference Optimization (DPO) for aligning Large Language Models (LLMs), yet existing methods rely on a one-dimensional view of difficulty. In this work, we reframe alignment difficulty as a two-dimensional space...

📖 Read original article


161. A Vision Toward Energy-Efficient Domain-Specific Artificial Intelligence Models and Agents ​

Author: Abhijit Chatterjee, Niraj K. Jha, Jonathan D. Cohen, Thomas L. Griffiths, Hongjing Lu, Diana Marculescu, Ashiqur Rasul, Wenrui Xu, Keshab K. Parhi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2510.22052v2 Announce Type: replace Abstract: The field of artificial intelligence (AI) has taken a tight hold on broad aspects of society, industry, business, and governance in ways that dictate the prosperity and might of the world's economies. The AI market size is projected to grow from {...

📖 Read original article


162. MetaHGNIE: Meta-Path Induced Hypergraph Contrastive Learning in Heterogeneous Knowledge Graphs ​

Author: Jiawen Chen, Yanyan He, Qi Shao, Mengli Wei, Duxin Chen, Wenwu Yu, Yanlong Zhao
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2512.12477v2 Announce Type: replace Abstract: Estimating node importance in heterogeneous knowledge graphs is a fundamental problem underlying recommendation, search, and knowledge decision systems. However, most existing methods rely on pairwise message passing mechanisms that fail to capture...

📖 Read original article


163. SimRPD: Optimizing Recruitment Proactive Dialogue Agents through Simulator-Based Data Evaluation and Selection ​

Author: Zhiyong Cao, Dunqiang Liu, Qi Dai, Haojun Xu, Huai Yuen Khor, Hao Wang, Huan He, Yafei Liu, Ke Ma, Ruqian Shi, Sicheng Zhou, Sijia Yao
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2601.02871v3 Announce Type: replace Abstract: Task-oriented proactive dialogue agents play a pivotal role in recruitment, particularly for steering conversations towards specific business outcomes, such as acquiring social-media contacts for private-channel conversion. Although supervised fine...

📖 Read original article


164. Conversational AI for Rapid Scientific Prototyping: A Case Study on ESA's ELOPE Competition ​

Author: Nils Einecke
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2601.04920v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as coding partners, yet their role in accelerating scientific discovery remains underexplored. This paper presents a case study of using ChatGPT for rapid prototyping in ESA's ELOPE (Event-based Lu...

📖 Read original article


165. RetailBench: Evaluating Long-Horizon Autonomous Decision-Making and Strategy Stability of LLM Agents in Realistic Retail Environments ​

Author: Linghua Zhang, Jun Wang, Jingtong Wu, Zhisong Zhang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2603.16453v3 Announce Type: replace Abstract: Large language model (LLM) agents have made rapid progress on short-horizon, well-scoped tasks, yet their ability to sustain coherent decisions in dynamic long-horizon environments remains uncertain. We introduce RetailBench, a data-grounded simula...

📖 Read original article


166. Retrieval-Augmented Generation Must Move Beyond Factual Grounding to Represent Diverse Opinions ​

Author: Aditya Agrawal, Alwarappan Nakkiran, Darshan Fofadiya, Alex Karlsson, Harsha Aduri, Aman Singh Thakur
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.IR

arXiv:2604.12138v3 Announce Type: replace Abstract: This position paper argues that Retrieval-Augmented Generation (RAG) systems exhibit a factual bias-optimizing for epistemic uncertainty reduction while ignoring the aleatoric uncertainty inherent in opinion-rich content. This misalignment demands ...

📖 Read original article


167. The Power of Power Law: Asymmetry Enables Compositional Reasoning ​

Author: Zixuan Wang, Xingyu Dang, Jason D. Lee, Kaifeng Lyu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2604.22951v2 Announce Type: replace Abstract: Natural language data follows a power-law distribution, with most knowledge and skills appearing at very low frequency. While a common intuition suggests that reweighting or curating data towards a uniform distribution may help models better learn ...

📖 Read original article


168. SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition ​

Author: Jayanta Dey, Shikhar Srivastava, Itamar Lerner, Christopher Kanan, Dhireesha Kudithipudi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.00732v3 Announce Type: replace Abstract: Learning long-range non-stationary temporal patterns remains a core challenge for modern sequence models, particularly in strict streaming settings. In these settings, data arrive sequentially and must be processed in a single pass without simultan...

📖 Read original article


169. Deployment-Time Memorization in Foundation-Model Agents ​

Author: Lei (Rachel), Chen, Guilin Zhang, Kai Zhao, Dalmo Cirne, Andy Olsen, Xu Chu, Zeke Miller, Alet Blanken, Amine Anoun, Jerry Ting
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.MA

arXiv:2606.10062v2 Announce Type: replace Abstract: Foundation-model agents are increasingly long-lived systems that remember users across interactions, making memorization an explicit deployment-time function rather than solely a property of model weights. Existing work addresses parametric memoriz...

📖 Read original article


170. LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis ​

Author: Minh-Ha Nguyen, Erica Gray, Bryce A. Schuler, Chih-Ting Yang, Rizwan Hamid, Lingyao Li, Siyuan Ma, Thomas A. Cassini, Cathy Shyr
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2606.16149v2 Announce Type: replace Abstract: Rare disease diagnosis involves interpreting clinical and genetic findings through complex diagnostic reasoning. We investigated whether this reasoning could be translated into a portable policy for guiding general-purpose large language models (LL...

📖 Read original article


171. TNODEV: Toolbox for Neural ODE Verification ​

Author: Abdelrahman Sayed Sayed, Pierre-Jean Meyer, Mohamed Ghazel
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SY, eess.SY, math.DS

arXiv:2606.16567v2 Announce Type: replace Abstract: Neural ordinary differential equations (neural ODE) gained attention in safety critical settings such as continuous-time controllers for cyber-physical systems and classifiers integrated into automated decision pipelines, raising the question wheth...

📖 Read original article


172. InvestPhilBench: A Multi-Layer Benchmark for Evaluating Large Language Model Procedural Reasoning in Expert Investment Philosophy ​

Author: Mingguang Chen, Bo Qu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.25984v2 Announce Type: replace Abstract: Large language models are increasingly deployed as investment research assistants, yet no benchmark tests whether they can accurately reconstruct and apply the specific procedural decision frameworks of expert investors. We introduce InvestPhilBenc...

📖 Read original article


173. Narration-of-Thought: Inference-Time Scaffolding for Defeasible Ethical Reasoning in Large Language Models ​

Author: Patrick Cooper, Alvaro Velasquez
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CY

arXiv:2606.26366v2 Announce Type: replace Abstract: Standard chain-of-thought on moral dilemmas exhibits two failure modes: stakeholder collapse (the trace names at most one party with a stake in the outcome) and uncertainty suppression (no explicit unknowns or hedges before committing to an action)...

📖 Read original article


174. Theoria: Rewrite-Acceptability Verification over Informal Reasoning States ​

Author: Michael Saldivar, Ben Slivinski
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.LO, cs.SE

arXiv:2607.01223v3 Announce Type: replace Abstract: When should an AI system's answer be trusted? Formal proof assistants offer certainty but cannot reach most of the problem distribution; scalar LLM judges offer coverage but produce opaque scores that cannot be audited after the fact and are subjec...

📖 Read original article


175. ContextSniper: AntTrail's Token-Efficient Code Memory for Repository-Level Program Repair ​

Author: Chiwang Luk, Matin Mohammad Najafi, Zhifeng Jia, Wei Yang, Xiuchang Li, Jinwei Zhu, Yang Ren, Lei Chen, Gao Cong
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.01916v3 Announce Type: replace Abstract: Large language model agents can repair real repository issues, but they often spend large context budgets on whole-file reads, broad searches, and long terminal outputs where useful evidence is mixed with irrelevant code and logs. This paper presen...

📖 Read original article


176. Distributed Attacks in Persistent-State AI Control ​

Author: Josh Hills, Ida Caspary, Asa Cooper Stickland
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI

arXiv:2607.02514v2 Announce Type: replace Abstract: As AI coding agents become more autonomous, they increasingly ship code iteratively, with the codebase persisting across sessions. This persistence creates a new attack surface: a misaligned or prompt-injected agent can distribute attacks across pu...

📖 Read original article


177. PolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents ​

Author: Hongliang Li, Yijin Liu, Zhiwei Zhang, Zihe Liu, Xinyue Lou, Jinan Xu, Fandong Meng, Kaiyu Huang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.06008v2 Announce Type: replace Abstract: Large language model (LLM) agents have shown strong performance in long-horizon tasks that require planning, tool use, and interaction with external environments. However, most existing benchmarks implicitly assume a monolingual setting, where the ...

📖 Read original article


178. From Application-Layer Simulation to Native Meta-Architecture: Structural Tension as an Endogenous Driver for Heterogeneous AI Evolution ​

Author: Heting Mao
Published: 7/11/2026, 4:00:00 AM
Categories: cs.AI, cs.CL

arXiv:2607.06269v2 Announce Type: replace Abstract: Current large language models (LLMs) are stateless across inference sessions: their behavior is fully determined by input at inference time, and any higher-order cognitive architecture must be simulated at the application layer through prompt engin...

📖 Read original article


179. Accurate Portraits of Scientific Resources and Knowledge Service Components ​

Author: Yue Wang, Zhe Xue, Ang Li
Published: 7/11/2026, 4:00:00 AM
Categories: cs.DL, cs.AI

arXiv:2204.04883v2 Announce Type: replace-cross Abstract: With the advent of the cloud computing era, the cost of creating, capturing, and managing information has gradually decreased. The amount of data on the Internet is showing explosive growth, and more scientific and technological resources are...

📖 Read original article


180. Knowledge Graph and Accurate Portrait Construction of Scientific and Technological Academic Conferences ​

Author: Runyu Yu, Zhe Xue, Ang Li
Published: 7/11/2026, 4:00:00 AM
Categories: cs.DL, cs.AI, cs.LG

arXiv:2204.04888v2 Announce Type: replace-cross Abstract: In recent years, with the continuous progress of science and technology, the number of scientific research achievements has increased rapidly. As an exchange platform and medium for scientific research achievements, scientific and technologic...

📖 Read original article


181. Retrieval of Scientific and Technological Resources for Experts and Scholars ​

Author: Suyu Ouyang, Yingxia Shao, Ang Li
Published: 7/11/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2204.06142v2 Announce Type: replace-cross Abstract: Institutions of higher learning, research institutes and other scientific research units have abundant scientific and technological resources of experts and scholars, and these talents with great scientific and technological innovation abilit...

📖 Read original article


182. The Contribution of XAI for the Safe Development and Certification of AI: An Expert-Based Analysis ​

Author: Benjamin Fresz, Vincent Philipp G"obels, Safa Omri, Danilo Brajovic, Andreas Aichele, Janika Kutz, Jens Neuh"uttler, Marco F. Huber
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CY, cs.AI

arXiv:2408.02379v2 Announce Type: replace-cross Abstract: Developing and certifying safe - or so-called trustworthy - AI has become an increasingly salient issue, especially in light of upcoming regulation such as the EU AI Act. In this context, the black-box nature of machine learning models limits...

📖 Read original article


183. ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation ​

Author: Pengcheng Huang, Zhenghao Liu, Yukun Yan, Haiyan Zhao, Xiaoyuan Yi, Hao Chen, Zhiyuan Liu, Maosong Sun, Tong Xiao, Ge Yu, Chenyan Xiong
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2502.15543v4 Announce Type: replace-cross Abstract: Large language models (LLMs) integrated with retrieval-augmented generation (RAG) have improved factuality by grounding outputs in external evidence. However, they remain susceptible to unfaithful generation, where outputs contradict retrieve...

📖 Read original article


184. Simulator Ensembles for Trustworthy Autonomous Driving Systems Testing ​

Author: Lev Sorokin, Matteo Biagiola, Andrea Stocco
Published: 7/11/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.RO

arXiv:2503.08936v3 Announce Type: replace-cross Abstract: Scenario-based testing with driving simulators is extensively used to identify failing conditions of automated driving assistance systems (ADAS). However, existing studies have shown that repeated test execution in the same as well as in dist...

📖 Read original article


185. Concept-as-Tree: A Controllable Synthetic Data Framework Makes Stronger Personalized VLMs ​

Author: Ruichuan An, Kai Zeng, Ming Lu, Sihan Yang, Renrui Zhang, Huitong Ji, Hao Liang, Wentao Zhang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2503.12999v4 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have demonstrated exceptional performance in various multi-modal tasks. Recently, there has been an increasing interest in improving the personalization capabilities of VLMs. To better integrate user-provided con...

📖 Read original article


186. ToDMA: Large Model-Driven Massive Token Communications for Semantic Multiple Access ​

Author: Li Qiao, Mahdi Boloursaz Mashhadi, Zhen Gao, Robert Schober, Deniz G"und"uz
Published: 7/11/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, eess.SP, math.IT

arXiv:2505.10946v3 Announce Type: replace-cross Abstract: Token communications (TokenCom) is an emerging generative semantic communication paradigm, where tokens serve as compact representation units across modalities. Their contextual dependencies can be exploited by pretrained large models for sem...

📖 Read original article


187. Less Is More: Reducing Token Counts Without Compromising Performance ​

Author: Gyeongje Cho, Yeonkyoung So, Sangmin Lee, Jaejin Lee
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2506.15138v2 Announce Type: replace-cross Abstract: Tokenization directly affects the inference efficiency of large language models, since fragmented tokenization increases sequence length and generation cost. Although longer, multi-word tokens can reduce fertility, naively adding them often d...

📖 Read original article


188. How Causal Abstraction Underpins Computational Explanation ​

Author: Atticus Geiger, Jacqueline Harding, Thomas Icard
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2508.11214v2 Announce Type: replace-cross Abstract: Explanations of cognitive behavior often appeal to computations over representations. What does it take for a system to implement a given computation over suitable representational vehicles within that system? We argue that the language of ca...

📖 Read original article


189. TOPO-Bench: An Open-Source Topological Mapping Evaluation Framework with Quantifiable Perceptual Aliasing ​

Author: Jiaming Wang, Jizhuo Chen, Diwen Liu, Harold Soh
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2510.04100v2 Announce Type: replace-cross Abstract: Topological mapping offers a compact and robust representation for navigation, but progress in the field is hindered by the lack of standardized evaluation metrics, datasets, and protocols. Existing systems are assessed using different enviro...

📖 Read original article


190. MultiFair: Multimodal Balanced Fairness-Aware Medical Classification with Dual-Level Gradient Modulation ​

Author: Md Zubair, Hao Zheng, Grayson W. Armstrong, Lucy Q. Shen, Gabriela Wilson, Yu Tian, Xingquan Zhu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.CY

arXiv:2510.07328v2 Announce Type: replace-cross Abstract: Medical decision systems increasingly rely on data from multiple sources to ensure reliable and unbiased diagnosis. However, existing multimodal learning models fail to achieve this goal because they often overlook two critical challenges. Fi...

📖 Read original article


191. Adaptive Generation of Bias-Eliciting Questions for LLMs ​

Author: Robin Staab, Jasper Dekoninck, Maximilian Baader, Martin Vechev
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CY, cs.AI

arXiv:2510.12857v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now widely deployed in user-facing applications, reaching hundreds of millions of users worldwide. Despite their widespread adoption, growing reliance on their outputs raises significant concerns, particularly...

📖 Read original article


192. OREN: Octree Residual Network for Real-Time Euclidean Signed Distance Mapping ​

Author: Zhirui Dai, Qihao Qian, Tianxing Fan, Nikolay Atanasov
Published: 7/11/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV

arXiv:2510.18999v3 Announce Type: replace-cross Abstract: Reconstructing signed distance functions (SDFs) from point cloud data benefits many robot autonomy capabilities, including localization, mapping, motion planning, and control. Methods that support online and large-scale SDF reconstruction oft...

📖 Read original article


193. MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection ​

Author: Leena Alghamdi, Muhammad Usman, Hafeez Anwar, Abdul Bais, Saeed Anwar
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, eess.IV

arXiv:2511.12810v2 Announce Type: replace-cross Abstract: Camouflaged object detection is an emerging and challenging computer vision task that requires identifying and segmenting objects that blend seamlessly into their environments due to high similarity in color, texture, and size. This task is f...

📖 Read original article


194. Deep Neural Networks as Discrete Dynamical Systems: Implications for Physics-Informed Learning ​

Author: Abhisek Ganguly, Santosh Ansumali, Sauro Succi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.00473v4 Announce Type: replace-cross Abstract: We revisit the analogy between feed-forward deep neural networks (DNNs) and discrete dynamical systems derived from neural integral equations and their corresponding partial differential equation (PDE) forms. A comparative analysis between th...

📖 Read original article


195. CoCo-Fed: A Unified Framework for Memory- and Communication-Efficient Federated Learning at the Wireless Edge ​

Author: Zhiheng Guo, Zhaoyang Liu, Zihan Cen, Chenyuan Feng, Xinghua Sun, Xiang Chen, Tony Q. S. Quek, Xijun Wang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, math.IT

arXiv:2601.00549v2 Announce Type: replace-cross Abstract: The deployment of large-scale neural networks within the Open Radio Access Network (O-RAN) architecture is pivotal for enabling native edge intelligence. However, this paradigm faces two critical bottlenecks: the prohibitive memory footprint ...

📖 Read original article


196. V-VLAPS: Value-Guided Planning for Vision-Language-Action Models ​

Author: Ke Ren, Ali Salamatian, Kieran Pattison, Cyrus Neary
Published: 7/11/2026, 4:00:00 AM
Categories: cs.RO, cs.AI

arXiv:2601.00969v3 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models provide strong action priors for robotic manipulation, but their reactive behavior can fail under distribution shift and long-horizon task structure. Recent VLA-guided planning methods improve execution by ...

📖 Read original article


197. SwinIFS: Landmark Guided Swin Transformer For Identity Preserving Face Super Resolution ​

Author: Habiba Kausar, Saeed Anwar, Omar Jamal Hammad, Abdul Bais
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2601.01406v2 Announce Type: replace-cross Abstract: Face super-resolution aims to recover high-quality facial images from severely degraded low-resolution inputs, but remains challenging due to the loss of fine structural details and identity-specific features. This work introduces SwinIFS, a ...

📖 Read original article


198. Bridging Cognitive Neuroscience and Graph Intelligence: Hippocampus-Inspired Multi-View Hypergraph Learning for Web Finance Fraud ​

Author: Rongkun Cui, Nana Zhang, Kun Zhu, Qi Zhang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.11073v3 Announce Type: replace-cross Abstract: Online financial services constitute an essential component of contemporary web ecosystems, yet their openness introduces substantial exposure to fraud that harms vulnerable users and weakens trust in digital finance. Such threats have become...

📖 Read original article


199. GenDA: Generative Data Assimilation on Complex Urban Areas via Classifier-Free Diffusion Guidance ​

Author: Francisco Giral, 'Alvaro Manzano, Ignacio G'omez, Ricardo Vinuesa, Soledad Le Clainche
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE

arXiv:2601.11440v3 Announce Type: replace-cross Abstract: Urban wind flow reconstruction is essential for assessing air quality, heat dispersion, and pedestrian comfort, yet remains challenging when only sparse sensor data are available. We propose GenDA, a generative data assimilation framework tha...

📖 Read original article


200. XFACTORS: Disentangled Information Bottleneck via Contrastive Supervision ​

Author: Alexandre Myara, Nicolas Bourriez, Thomas Boyer, Thomas Lemercier, Ihab Bendidi, Auguste Genovesio
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.21688v2 Announce Type: replace-cross Abstract: Disentangled representation learning aims to map independent factors of variation to independent representation components. On one hand, purely unsupervised approaches have proven successful on fully disentangled synthetic data, but fail to r...

📖 Read original article


201. Rethinking LLM-as-a-Judge: Representation-as-a-Judge with Small Language Models via Semantic Capacity Asymmetry ​

Author: Zhuochun Li, Yong Zhang, Ming Li, Yuelyu Ji, Yiming Zeng, Ning Cheng, Yun Zhu, Yanmeng Wang, Shaojun Wang, Jing Xiao, Daqing He
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2601.22588v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are widely used as reference-free evaluators via prompting, but this "LLM-as-a-Judge" paradigm is costly, opaque, and sensitive to prompt design. In this work, we investigate whether smaller models can serve as ef...

📖 Read original article


202. Towards Isolated Interventions via Almost Orthogonal Features in Language Models ​

Author: Moritz Miller, Florent Draye, Bernhard Sch"olkopf
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2602.04718v2 Announce Type: replace-cross Abstract: A central premise in mechanistic interpretability is that meaningful concepts in language models are represented by linear features in activation space. For such features to support reliable interventions, manipulating one feature should not ...

📖 Read original article


203. Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback ​

Author: Sein Kim, Sangwu Park, Hongseok Kang, Wonjoong Kim, Jimin Seo, Yeonjun In, Kanghoon Yoon, Hyunsik Jeon, Chanyoung Park
Published: 7/11/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2602.12612v2 Announce Type: replace-cross Abstract: Traditional methods for automating recommender system design, such as Neural Architecture Search (NAS), are often constrained by a fixed search space defined by human priors, limiting innovation to pre-defined operators. While recent LLM-driv...

📖 Read original article


204. An Online Reference-Free Evaluation Framework for Flowchart Image-to-Code Generation ​

Author: Giang Son Nguyen, Zi Pong Lim, Sarthak Ketanbhai Modi, Yon Shin Teo, Wenya Wang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL

arXiv:2602.13376v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly used in document processing pipelines to convert flowchart images into structured code (e.g., Mermaid). In production, these systems process arbitrary inputs for which no ground-truth code exists...

📖 Read original article


205. Curriculum Learning for Efficient Chain-of-Thought Distillation via Structure-Aware Masking and GRPO ​

Author: Bowen Yu, Maolin Wang, Sheng Zhang, Binhao Wang, Yi Wen, Jingtong Gao, Bowen Liu, Zimo Zhao, Wanyu Wang, Xiangyu Zhao
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17686v4 Announce Type: replace-cross Abstract: Distilling Chain-of-Thought (CoT) reasoning from large language models into compact student models presents a fundamental challenge: teacher rationales are often too verbose for smaller models to faithfully reproduce. Existing approaches eith...

📖 Read original article


206. Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization ​

Author: Theophilus Amaefuna, Hitesh Vaidya, Anshuman Chhabra, Ankur Mali
Published: 7/11/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, math.IT

arXiv:2603.00910v2 Announce Type: replace-cross Abstract: Layer-wise capacity in large language models is highly non-uniform: some layers contribute disproportionately to loss reduction, whereas others are nearly redundant. Existing layer-scoring methods provide sensitivity estimates but do not give...

📖 Read original article


207. Robust Weighted Triangulation of Causal Effects Under Model Uncertainty ​

Author: Rohit Bhattacharya, Ina Ocelli, Ted Westling
Published: 7/11/2026, 4:00:00 AM
Categories: stat.ME, cs.AI

arXiv:2603.01119v2 Announce Type: replace-cross Abstract: A fundamental challenge in causal inference with observational data is correct specification of a causal model. When there is model uncertainty, analysts may seek to use estimates from multiple candidate models that rely on distinct, and poss...

📖 Read original article


208. Human-like Object Grouping in Self-supervised Vision Transformers ​

Author: Hossein Adeli, Seoyoung Ahn, Andrew Luo, Mengmi Zhang, Nikolaus Kriegeskorte, Gregory Zelinsky
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, q-bio.NC

arXiv:2603.13994v3 Announce Type: replace-cross Abstract: Vision foundation models trained with self-supervised objectives achieve strong performance across diverse tasks and exhibit emergent object segmentation properties. However, their alignment with human object perception remains poorly underst...

📖 Read original article


209. Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies ​

Author: Mumuksh Tayal, Manan Tayal, Ravi Prakash
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.15136v2 Announce Type: replace-cross Abstract: Offline safe reinforcement learning (RL) seeks reward-maximizing policies from static datasets under strict safety constraints. Existing methods often rely on soft expected-cost objectives or iterative generative inference, which can be insuf...

📖 Read original article


210. PhasorFlow: A Python Library for Unit Circle Based Computing ​

Author: Dibakar Sigdel, Namuna Panday
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.15886v3 Announce Type: replace-cross Abstract: We present PhasorFlow, an open-source Python library for computing on the $S^1$ unit circle. Inputs are encoded as complex phasors $z=e^{i\phi}$ on the $N$-torus ($\mathbb{T}^N$); as computation proceeds through unitary wave-interference gate...

📖 Read original article


211. The Phasor Transformer: Resolving Attention Bottlenecks on the Unit Circle ​

Author: Dibakar Sigdel
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.17433v2 Announce Type: replace-cross Abstract: Transformer models have redefined sequence learning, yet dot-product self-attention introduces a quadratic token-mixing bottleneck for long-context time-series. We introduce the Phasor Transformer block, a phase-native alternative representin...

📖 Read original article


212. StateLinFormer: Stateful Training Enhancing Long-term Memory in Navigation ​

Author: Zhiyuan Chen, Yuxuan Zhong, Fan Wang, Bo Yu, Pengtao Shao, Shaoshan Liu, Ning Ding
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2603.23571v2 Announce Type: replace-cross Abstract: Effective navigation intelligence relies on long-term memory to support both immediate generalization and sustained adaptation. However, existing approaches face a dilemma: modular systems rely on explicit mapping but lack flexibility, while ...

📖 Read original article


213. Echoes: A semantically-aligned music deepfake detection dataset ​

Author: Octavian Pascu, Dan Oneata, Horia Cucu, Nicolas M. Muller
Published: 7/11/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, eess.AS

arXiv:2603.23667v2 Announce Type: replace-cross Abstract: We introduce Echoes, a new dataset for music deepfake detection designed for training and benchmarking detectors under realistic and provider-diverse conditions. Echoes comprises 4,468 tracks (131 hours of audio) spanning multiple genres (pop...

📖 Read original article


214. Neural Harmonic Textures for High-Quality Primitive Based Neural Reconstruction ​

Author: Jorge Condor, Nicolas Moenne-Loccoz, Merlin Nimier-David, Piotr Didyk, Zan Gojcic, Qi Wu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.GR, cs.LG

arXiv:2604.01204v3 Announce Type: replace-cross Abstract: Primitive-based methods such as 3D Gaussian Splatting have recently become the state-of-the-art for novel-view synthesis and related reconstruction tasks. Compared to neural fields, these representations are more flexible, adaptive, and scale...

📖 Read original article


215. LSRM: High-Fidelity Object-Centric Reconstruction via Scaled Context Windows ​

Author: Zhengqin Li, Cheng Zhang, Jakob Engel, Zhao Dong
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2604.05182v2 Announce Type: replace-cross Abstract: We introduce the Large Sparse Reconstruction Model to study how scaling transformer context windows affects feed-forward 3D reconstruction. Although recent object-centric feed-forward methods produce robust, high-quality reconstructions, they...

📖 Read original article


216. Persona Matters: Effects of Activation Steering on Short Answer Generation and Scoring ​

Author: Yongchao Wu, Aron Henriksson
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2604.07102v2 Announce Type: replace-cross Abstract: Activation-based steering enables inference-time personalization of large language models, but its effects in educational applications are not well understood. We study activation-based persona vectors representing seven character traits in s...

📖 Read original article


217. Beyond Attention Scores: SVD-Based Vision Token Pruning for Efficient Vision-Language Models ​

Author: Yvon Apedo, Martyna Poreba, Michal Szczepanski, Samia Bouchafa
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2604.11530v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have revolutionized multi-modal learning by jointly processing visual and textual information. Yet, they face significant challenges due to the high computational and memory demands of processing long sequences o...

📖 Read original article


218. Peer-Predictive Self-Training for Language Model Reasoning ​

Author: Shi Feng, Hanlin Zhang, Fan Nie, Sham Kakade, Yiling Chen
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.GT

arXiv:2604.13356v3 Announce Type: replace-cross Abstract: Mechanisms for continued self-improvement of language models without external supervision remain an open challenge. We propose Peer-Predictive Self-Training (PST), a label-free fine-tuning framework in which multiple language models improve c...

📖 Read original article


219. Predicting Scale-Up of Metal-Organic Framework Syntheses with Large Language Models ​

Author: Peter Walther, Hongrui Sheng, Xinxin Liu, Bin Feng, Reid Coyle, Xinhua Yan, Kyle Smith, Harrison Kayal, Shyam Chand Pal, Zhiling Zheng
Published: 7/11/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.AI

arXiv:2604.20899v2 Announce Type: replace-cross Abstract: Scalable synthesis remains the gate between MOF discovery and industrial deployment, as scale-up know-how is fragmented across disparate reports. We introduce ScaleMOF, a literature-mined dataset and a positive-unlabeled learning strategy tha...

📖 Read original article


220. DeepTutor: Towards Agentic Personalized Tutoring ​

Author: Bingxi Zhao, Jiahao Zhang, Xubin Ren, Zirui Guo, Tianzhe Chu, Yi Ma, Chao Huang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CL

arXiv:2604.26962v3 Announce Type: replace-cross Abstract: Education is one of the most promising real-world applications for Large Language Models (LLMs). However, current LLMs rely on static pre-training knowledge and lack adaptation to individual learners, while existing RAG systems fall short in ...

📖 Read original article


221. LoKA: Low-precision Kernel Applications for Recommendation Models At Scale ​

Author: Liang Luo, Yinbin Ma, Quanyu Zhu, Vasiliy Kuznetsov, Yuxin Chen, Neng Shi, Jian Jiao, Jiecao Yu, Buyun Zhang, Tongyi Tang, Xiaohan Wei, Yanli Zhao, Zeliang Chen, Yuchen Hao, Venkatesh Ranganathan, Sandeep Parab, Yantao Yao, Maxim Naumov, Chunzhi Yang, Shen Li, Ellie Wen, Wenlin Chen, Santanu Kolay, Chunqiang Tang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.10886v3 Announce Type: replace-cross Abstract: Recent GPU generations deliver significantly higher FLOPs using lower-precision arithmetic, such as FP8. While successfully applied to large language models (LLMs), its adoption in large recommendation models (LRMs) has been limited. This is ...

📖 Read original article


222. Second-Order Actor-Critic Methods for Discounted MDPs via Policy Hessian Decomposition ​

Author: Sanjeev Manivannan, Shuban V
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.14982v2 Announce Type: replace-cross Abstract: We address the discounted reward setting in reinforcement learning (RL). To mitigate the value approximation challenges in policy gradient methods, actor-critic approaches have been developed and are known to converge to stationary points und...

📖 Read original article


223. MasFACT: Continual Multi-Agent Topology Learning via Geometry-Aware Posterior Transfer ​

Author: Xuefei Wang, Jialu Wang, Fengbo Zhang, Yihan Hu, Di Zhang, Yutong Ye, Yikun Ban, Jun Han, Ruijie Wang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.17361v2 Announce Type: replace-cross Abstract: Multi-agent systems (MAS) powered by large language models (LLMs) have emerged as a powerful paradigm for complex problem solving, where performance critically depends on the underlying inter-agent communication topology. However, existing to...

📖 Read original article


224. CriterAlign: Criterion-Centric Rationale Alignment for Code Preference Judging ​

Author: Zhenyu Li, Aleksandar Cvejic, Zehui Chen, Peter Wonka
Published: 7/11/2026, 4:00:00 AM
Categories: cs.SE, cs.AI

arXiv:2605.19665v2 Announce Type: replace-cross Abstract: Pairwise human preference prediction is central to evaluating code-generation systems, where quality often depends on task-specific trade-offs beyond functional correctness. While rubric-based LLM judges improve interpretability by decomposin...

📖 Read original article


225. MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks ​

Author: Han Zhang, Wanting Jiang, Tomasz Kornuta, Tian Zheng, Vidya Murali
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2605.21917v2 Announce Type: replace-cross Abstract: Training Vision Language Models (VLMs) for video event reasoning requires high-quality structured annotations capturing not only what happened, but when, where, why, and with what consequence, at a scale manual labelling cannot support. We pr...

📖 Read original article


226. Correcting Visual Blur Induced by Attention Distraction to Reduce Hallucinations: Algorithm and Theory ​

Author: Quanjiang Li, Zhiming Liu, Wei Luo, Tingjin Luo, Chenping Hou
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2605.24602v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) frequently suffer from object hallucinations, yet the visual perceptual mechanism underlying this failure remains poorly understood. In this work, we reveal that hallucinations are strongly associated ...

📖 Read original article


227. When RLHF Fails: A Mechanistic Taxonomy of Reward Hacking, Collapse, and Evaluator Gaming ​

Author: Zelalem Abahana, David Evans, Satish Mahadevan Srinivasan, Matjaz Gams
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.03238v2 Announce Type: replace-cross Abstract: RLHF evaluation should track how failures emerge, where they localize, and which warning signals appear before external quality degrades. We study this problem with a compact RLHF pipeline built for this paper, including PPO, DPO, uncertainty...

📖 Read original article


228. Autonomous heterogeneous catalyst discovery with a self-evolving multi-agent digital twin ​

Author: Zhilong Song, Zongmin Zhang, Lixue Cheng
Published: 7/11/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.AI, physics.chem-ph

arXiv:2606.05050v2 Announce Type: replace-cross Abstract: Theoretical heterogeneous catalysis promises rapid catalyst discovery, yet computational and machine-learning predictions often deviate from experiment and stay confined to narrow material families, for want of a faithful, condition-aware cat...

📖 Read original article


229. Temporal Preference Concepts and their Functions in a Large Language Model ​

Author: Ian Rios-Sialer, Shantanu Darveshi, Shuai Jiang, Avigya Paudel, Anastasiia Pronina, Ipshita Bandyopadhyay, Justin Shenk
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2606.05194v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly being deployed to make decisions that require trading off near-term gains against long-term consequences, yet little is known about how they internally represent or resolve these tradeoffs. In thi...

📖 Read original article


230. EasyLens: A Training-Free Plug-and-Play Subtle-Lesion Representation Amplifier for Medical Vision-Language Models ​

Author: Qiwei Zeng, Hao Wang, Jinghao Lin, Shuchang Ye, Yuezhe Yang, Yige Peng, Haoyuan Che, Jinman Kim, Lei Bi
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2606.06379v2 Announce Type: replace-cross Abstract: Medical vision-language models (VLMs) have shown increasing potential for clinical image interpretation, including lesion detection and report generation. However, their practical utility remains limited by insufficient sensitivity to subtle ...

📖 Read original article


231. MeCo: One-Step MeanFlow-based Corrector for Multi-Channel Speech Separation ​

Author: Dohwan Kim, Jung-Woo Choi
Published: 7/11/2026, 4:00:00 AM
Categories: eess.AS, cs.AI

arXiv:2606.09677v3 Announce Type: replace-cross Abstract: While discriminative models for multi-channel speech separation excel in reference-based metrics, they often exhibit suboptimal human listening quality. To address this, we propose a novel MeanFlow-based one-step generative corrector (MeCo). ...

📖 Read original article


232. KG-SoftMAP: Soft Knowledge-Graph Priors for Bayesian Network Structure Learning from Sparse Discrete Data ​

Author: Guoliang Xu, James E. Corter
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.10358v3 Announce Type: replace-cross Abstract: Learning Bayesian network (BN) structure from sparse discrete data is hard: when each instance records only a few variables, most variable pairs lack the joint observations needed for reliable scoring, and data-only methods recover little str...

📖 Read original article


233. Bridging Modal Isolation in Interleaved Thinking: Supervising Modality Transitions via Stepwise Reinforcement ​

Author: Tingyu Li, Le Zhou, Siyuan Li, Yujun Wu, Xinglong Xu, Jingxuan Wei, Conghui He, Cheng Tan
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2606.12886v2 Announce Type: replace-cross Abstract: Interleaved thinking, where a unified multimodal model alternates between textual reasoning and visual generation, has shown promise on spatial and physical tasks. However, in complex long-chain scenarios, we identify a fundamental failure mo...

📖 Read original article


234. Training and Evaluating Diffusion Policies with Long Context Lengths ​

Author: Abhinav Agarwal, Adam Wei, Taylan Kargin, Michael Zeng, Cole Becker, Arif Kerem Dayi, Pablo Parrilo, Asuman Ozdaglar, Russ Tedrake
Published: 7/11/2026, 4:00:00 AM
Categories: cs.RO, cs.AI

arXiv:2606.16447v2 Announce Type: replace-cross Abstract: Imitation learning has enabled highly-dexterous robotic manipulation from RGB observations. Policies trained with these methods, however, typically condition robot actions on only a short history of observations. These policies cannot solve t...

📖 Read original article


235. Hierarchical Control in Multi-Agent Games: LLM-based Planning and RL Execution ​

Author: Jannik H"osch, Alessandro Sestini, Florian Fuchs, Amir Baghi, Joakim Bergdahl, Iolanda Leite, Konrad Tollmar, Jean-Philippe Barrette-LaPierre, Linus Gissl'en
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.20014v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has achieved strong performance in sequential decision-making, yet scaling to complex multi-agent environments remains challenging due to sparse rewards, large state-action spaces, and the difficulty of learning co...

📖 Read original article


236. Does Mixture-of-Experts Actually Help Inference on Consumer and Edge Hardware? An Empirical Study ​

Author: Alfarizy Alfarizy, Hung Truong Thanh Nguyen, Ren'e Richard, Roozbeh Razavi-Far, Hung Cao
Published: 7/11/2026, 4:00:00 AM
Categories: cs.PF, cs.AI

arXiv:2606.21428v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) language models are often described as ideal for resource-constrained inference. Each token activates only a small subset of experts, so the per-token compute cost, in floating-point operations (FLOPs), resembles that...

📖 Read original article


237. Improving Engine Sound Analysis in Hot-Test Environments via a RAB-U-Net (Residual Attention Block U-Net) Noise Removal Method ​

Author: Raheleh Mohseni, Mahdi Aliyari Shoorehdeli
Published: 7/11/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, eess.AS

arXiv:2606.21887v2 Announce Type: replace-cross Abstract: During hot tests on a production line, engine-sound analysis is crucial to ensuring product quality and performance. However, background noise often interferes with accurate sound analysis, leading to potential errors in engine diagnostics. T...

📖 Read original article


238. video-SALMONN-R$^3$: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding ​

Author: Yixuan Li, Guangzhi Sun, Yudong Yang, Chao Zhang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.SD

arXiv:2606.24477v2 Announce Type: replace-cross Abstract: Video large language models (LLMs) are often constrained by computation and memory budgets, leading them to use reduced frame rates and spatial resolutions, which may cause them to miss critical information for question answering (QA). A prac...

📖 Read original article


239. Play2Perfect: What Matters in Dexterous Play Pretraining for Precise Assembly? ​

Author: Tyler Ga Wei Lum, Kushal Kedia, C. Karen Liu, Jeannette Bohg
Published: 7/11/2026, 4:00:00 AM
Categories: cs.RO, cs.AI

arXiv:2606.26428v2 Announce Type: replace-cross Abstract: Multi-fingered robots promise the speed and dexterity of human hands, yet challenging problems such as precise assembly have remained out of reach. These tasks are contact-rich, making data collection for imitation learning difficult, and spa...

📖 Read original article


240. How to Leverage Synthetic Speech for LLM-Based ASR Systems? ​

Author: Yanis Labrak, Dairazalia Sanchez-Cortes, Sergio Burdisso, S'everin Baroudi, Shashi Kumar, Esa'u Villatoro-Tello, Srikanth Madikeri, Manjunath K E, Old\v{r}ich Plchot, Kadri Hacio\u{g}lu, Petr Motlicek, Andreas Stolcke
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2606.29031v2 Announce Type: replace-cross Abstract: In regulated domains such as banking and healthcare, where privacy constraints make real speech costly to collect and retain, synthetic speech from modern text-to-speech (TTS) is an appealing alternative for training automatic speech recognit...

📖 Read original article


241. A Stochastic--Geometric Theory of Scaling Laws in Grokking ​

Author: R'ois'in Luo, Christian Gagn'e, Jonas Ngnaw'e, Ihsan Ullah, Karyn Morrissey
Published: 7/11/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2606.30388v3 Announce Type: replace-cross Abstract: Delayed generalization (\ie~grokking) refers to the phenomenon in which a neural network fits its training data early in training but only begins to generalize after a prolonged delay, often through an abrupt transition. Despite extensive emp...

📖 Read original article


242. Diffusion-GR2: Diffusion Generative Reasoning Re-ranker ​

Author: Zhuoxuan Zhang (Yang), Kangqi Ni (Yang), Yuhang Chen (Yang), Mingfu Liang (Yang), Xiaohan Wei (Yang), Yunchen Pu (Yang), Fei Tian (Yang), Chonglin Sun (Yang), Frank Shyu (Yang), Adam (Yang), Song, Sandeep Pandey, Luke Simon, Tianlong Chen, Xi Liu
Published: 7/11/2026, 4:00:00 AM
Categories: cs.IR, cs.AI

arXiv:2607.01170v3 Announce Type: replace-cross Abstract: Generative reasoning re-rankers achieve strong recommendation accuracy by emitting a chain-of-thought before re-ordering a candidate list, but they are slow at inference: an autoregressive (AR) decoder spends one sequential forward pass per r...

📖 Read original article


243. Generalization in offline RL: The structure is more important than the amount of pessimism ​

Author: Max Weltevrede, Matthijs T. J. Spaan, Wendelin B"ohmer
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.02288v2 Announce Type: replace-cross Abstract: While pessimism counteracts overestimation bias in offline reinforcement learning (RL), being overly conservative has been associated with hindering certain forms of generalization. However, in this paper we demonstrate that being overly pess...

📖 Read original article


244. GAP-GDRNet: Geometry-aware monocular 6D pose estimation for spacecraft using synthetic geometric supervision ​

Author: Zongwu Xie, Yonglong Zhang, Yifan Yang, Yang Liu, Guanghu Xie
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.02360v3 Announce Type: replace-cross Abstract: Monocular spacecraft 6D pose estimation remains difficult under weak texture, thin structures, illumination variation, and occlusion. This article presents GAP-GDRNet, a geometry-aware RGB framework built on GDR-Net for a single-target synthe...

📖 Read original article


245. TAG: A Lightweight Framework for Test-Driven Agentic Artifact Generation ​

Author: Yaniv Melamed, Yoni Zukerman, Michal Shechter, Miri Weissler, Ashwin Patil, Hani Neuvirth-Telem
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.SE

arXiv:2607.02615v2 Announce Type: replace-cross Abstract: Generating structured artifacts with Large Language Models - e.g.\ database queries, threat framework mappings, entity schemas - is relatively straightforward; however, making them reliable enough for production deployments presents challenge...

📖 Read original article


246. Where do LLMs Fall Short in CBT-Guided Affective Reasoning? ​

Author: Vaishnavi Sinha, Pooja Guttal, Pranay Deep Reddy Katike, Vishal Sinha, Gerald Ndawula, Lira Yoon, Andrea Kleinsmith, Manas Gaur
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC, cs.IR

arXiv:2607.02885v2 Announce Type: replace-cross Abstract: Cognitive Behavioral Therapy (CBT) provides a structured framework for understanding a user's mental state by examining the interaction between cognitive and behavioral factors. However, out-of-the-box LLMs respond fluently and empathetically...

📖 Read original article


247. MambaLIE: Scene Light Intensity-Boosted Low-Light Image Enhancement with State Space Model ​

Author: Wanshu Fan, Xiangyu Li, Cong Wang, Kin-man Lam, Xin Yang, Haiyan Zhang, Dongsheng Zhou
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.03013v2 Announce Type: replace-cross Abstract: Images captured by consumer electronic devices, such as mobile phones and digital cameras, often suffer from low-light degradation due to sensor limitations and imaging pipelines, which degrades visual quality and affects downstream vision ta...

📖 Read original article


248. Phase-Preserving Trimodal Transformer for Tropical Forest Biomass Estimation Using Optical and PolInSAR Data ​

Author: Luiz Felipe Parente Santiago (Instituto de Computa\c{c}~ao, Universidade Federal do Amazonas, Instituto de Pesquisas do Ex'ercito na Amaz^onia), Felipe Ferrari (Instituto Militar de Engenharia), Daniel Rodrigues dos Santos (Instituto Militar de Engenharia), Rosiane de Freitas (Instituto de Computa\c{c}~ao, Universidade Federal do Amazonas)
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.03663v3 Announce Type: replace-cross Abstract: The accurate estimation of Above-Ground Biomass (AGB) in mature tropical forests remains a critical challenge in remote sensing, primarily due to the saturation of Synthetic Aperture Radar (SAR) signals in high-density areas and persistent cl...

📖 Read original article


249. LLM for the development of FCM ​

Author: Alexis Kafantaris
Published: 7/11/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.ET, cs.GL, cs.LG

arXiv:2607.04983v2 Announce Type: replace-cross Abstract: This article is about the development of a fuzzy cognitive map using a local large language model. In the light of recent advances it is evident that large language models, and even local large language models are capable of extracting quanti...

📖 Read original article


250. KVpop -- Key-Value Cache Compression with Predictive Online Pruning ​

Author: Lukas Hauzenberger, Niklas Schmidinger, Anamaria-Roberta Hartl, David Stap, Thomas Schmied, Sebastian B"ock, G"unter Klambauer, Sepp Hochreiter
Published: 7/11/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.05061v2 Announce Type: replace-cross Abstract: Key-value (KV) cache growth is a major bottleneck in autoregressive decoding, as memory and bandwidth scale linearly with context length. Existing KV eviction methods often rely on static heuristics or proxy scores, which poorly track future ...

📖 Read original article


251. Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation ​

Author: Haozhe Wang, Weijia Feng, Jinpeng Yu, Che Liu, Ping Nie, Fangzhen Lin, Jiaming Liu, Ruihua Huang, Jimmy Lin, Wenhu Chen, Cong Wei
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CV, cs.AI

arXiv:2607.05382v3 Announce Type: replace-cross Abstract: Visual generators excel at rendering, but they confidently fabricate what they do not know. User requests are unbounded, evolving, and deeply long-tailed: new characters, trending entities, post-cutoff events, and more. This world-knowledge b...

📖 Read original article


252. ResonatorLM: Causal Resonant Field Mixing for Efficient Long-Context Language Modeling ​

Author: Archie Chaudhury
Published: 7/11/2026, 4:00:00 AM
Categories: cs.CL, cs.AI

arXiv:2607.05583v2 Announce Type: replace-cross Abstract: Contemporary language models are dominated by the transformer architecture, which leverages self-attention mechanisms to enable more efficient, parallelized training across a wide set of documents and corpora. This has allowed transformers to...

📖 Read original article


253. Behavior Foundations for Quadruped Robots: ABot-C0 Technical Report ​

Author: Xufeng Zhao, Fuzhi Yang, Jianhui Chen, Li Gao, Zhang Meng, Jie Gao, Yao Zheng, Congyang Zhao, Tianxiong Lv, Menglin Yang, Minqi Gu, Yaru Zhao, Wenyu Liu, Honglin Han, Shihui Su, Zixiao Tang, Liu Liu, Mu Xu, Yang Cai, Wenbin Tang
Published: 7/11/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.HC, cs.LG

arXiv:2607.07370v2 Announce Type: replace-cross Abstract: The motion controller is one of the most fundamental modules in embodied intelligence systems. Driven by large-scale human motion-capture data and the motion-tracking paradigm, humanoid control has achieved remarkable progress in recent years...

📖 Read original article