arXiv cs.LG - 2026-07-30 ​
237 items collected.
1. Emergent Sparsity in Frozen Random CNN Feature Extractors for Deep Reinforcement Learning ​
Author: Scott M. Norton
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2607.26059v1 Announce Type: new Abstract: We report a striking phenomenon: deep reinforcement learning agents trained with frozen, randomly initialized CNN feature extractors spontaneously develop extremely sparse fully-connected representations, without any sparsity-inducing objective. In the...
2. Sim2Win: A Team-Agnostic, Event-Based Pre-Match Outcome Prediction and Tactical Profiling System for Football ​
Author: Mouad Zemzoumi, Amine Abouaomar
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26061v1 Announce Type: new Abstract: Pre-match tactical decision-making in professional football relies heavily on subjective expert analysis and identity-based scouting systems that cannot generalize to unseen teams. This paper presents Sim2Win, a team-agnostic, event-based pre-match tac...
3. Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback ​
Author: Yunpeng Chu
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.26094v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard approach for aligning large language models with human preferences, but its quality is limited by static, task-agnostic reward models. This mismatch leads to sparse learning signals and ...
4. Data Fusion and Contrastive Alignment for Unconstrained IR Molecular Structure Elucidation ​
Author: Ethan J. Mick, Campbell A. Sweet, Matthias J. Young, Derek T. Anderson
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26164v1 Announce Type: new Abstract: Automated molecular structure elucidation from infrared (IR) spectroscopy data has seen significant advancements in recent years, but its broad applicability is limited by a reliance on pre-determined chemical formulas provided as auxiliary model input...
5. Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models ​
Author: Anton de la Fuente, Arthur Conmy
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26173v1 Announce Type: new Abstract: Alignment training, model organisms, and toy models are usually treated as separate research areas. But projects in all three frequently use supervised fine-tuning (SFT) to pursue the same underlying goals. When projects share a goal, we should test wh...
6. Dynamic Parameterization Is Not Dynamic Inference ​
Author: Zongfei Li, Yuan-yih Shang, Guozhong Luo
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26192v1 Announce Type: new Abstract: Input-dependent controller coefficients are often treated as evidence of dynamic inference or computational savings. This interpretation conflates three properties: coefficient variation, dependence of a frozen model on how coefficients are assigned to...
7. Weak-to-Strong On-Policy Distillation ​
Author: Fangxu Yu, Zinan Lin, Xiaodong Liu, Weijia Xu, Michael Xu, Tianyi Zhou, Jianfeng Gao
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26246v1 Announce Type: new Abstract: On-policy distillation (OPD), which aligns a student with the teacher's token-level distribution on the student's own rollouts, is an effective paradigm for transferring capabilities across LLMs. Prevailing approaches assume a teacher at least as capab...
8. Between Gradient and Natural Gradient: A Continuum of LoRA Initializations ​
Author: Dianze Liu, Farshid Ghezelbash
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26247v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) fine-tunes large pretrained models at a fraction of the cost of full fine-tuning, but its performance depends strongly on how the adapters are initialized. Recent schemes initialize the adapters from the downstream loss gradi...
9. Early Verdicts, Better Budgets: Sequential Adaptive Rollout Allocation for Compute-Efficient RLVR ​
Author: Pixel Nomand, Elena Voss, Marcus Hale, Sofia Reyes
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26253v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is bottlenecked by rollout generation, yet many sampled prompts produce saturated groups (all responses correct or all incorrect) whose zero reward variance yields no policy-gradient signal. Existin...
10. Top-$k$ Pareto Bandits: Hypervolume Regret for Multi-Objective Slate Selection ​
Author: Nicolas Gutowski, Fabien Chhel, Alexandre Letard, Sylvain Lamprier
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.26273v1 Announce Type: new Abstract: We consider a stochastic multi-objective bandit problem where, at each round, the agent selects a slate of $k$ arms and observes their $d$-dimensional reward vectors under semi-bandit feedback. We do not aim at identifying a single optimal arm; instead...
11. FloDR: An invertible dimensionality reduction method based on a normalising flow ​
Author: Abdallah Baraka, Daniel Probst
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.HC
arXiv:2607.26278v1 Announce Type: new Abstract: It is common for two-dimensional embeddings of high-dimensional data to be read far beyond what they can support. Distances in and between clusters, the meaning behind empty spaces, and the amount of structure hidden at each point are generally invisib...
12. Entity Resolution in Practice: Lessons from a Self-Serve Pipeline ​
Author: Kaushik Pavani, Ganga Aluri, Pravin Jadhav, Neeraj Prasad, Kiran Sanka
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26298v1 Announce Type: new Abstract: We built and evaluated a self-serve entity resolution (ER) system on six benchmarks spanning 864 to 5M records, and three lessons emerged that are absent from existing ER literature. (1) No single matching algorithm wins everywhere - a self-serve pipel...
13. Learning Implicit Causal World Models from Multi-Agent Demonstrations ​
Author: Jasorsi Ghosh
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.MA, cs.RO
arXiv:2607.26336v1 Announce Type: new Abstract: In model-based reinforcement learning, world models exist as internal simulators, but their training often conflates statistical correlations with causal mechanisms. This problem is exacerbated in multi-agent systems where physical transitions are inte...
14. RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning ​
Author: Pushkal Kumar, Tucker Nielson, Tanish Kolhe, Shubham Zala, Vincent Li
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.IR
arXiv:2607.26339v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems ground large language models (LLMs) in external corpora, but this reliance exposes them to corpus poisoning: maliciously injected passages that manipulate retrieved evidence. We introduce RAGuard, a layered ...
15. Automorphism-Induced Non-Canonicity in Top-k Explanations of Graph Neural Networks ​
Author: Xin Xu, Siru Tao, Kaizhen Tan
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26344v1 Announce Type: new Abstract: A gradient-based GNN explainer given a molecule with two chemically equivalent nitro groups assigns them attribution scores that are equal to the last bit. It cannot do otherwise: message passing is exactly permutation equivariant, so any automorphism ...
16. MetaKoopman: Bayesian Meta-Learning of Koopman Operators for Modeling Structured Dynamics under Distribution Shifts ​
Author: Mahmoud Selim, Sriharsha Bhat, Karl H. Johansson
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.RO, cs.SY, eess.SY
arXiv:2607.26345v1 Announce Type: new Abstract: Modeling and forecasting nonlinear dynamics under distribution shifts is essential for robust decision-making in real-world systems. In this work, we propose MetaKoopman, a Bayesian meta-learning framework for modeling nonlinear dynamics through linear...
17. High-Order Markov Blanket Discovery via a k-Order Relaxation of the Faithfulness Assumption ​
Author: Loong Kuan Lee, Ragavi Krishnamoorthy, Nico Piatkowski
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.26357v1 Announce Type: new Abstract: The problem of learning the graphical Markov blanket (MB) of a variable from data has applications in many areas such as structure learning for Bayesian networks and Markov random fields, causal discovery, and feature selection. However, a common assum...
18. Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning ​
Author: Keegan Harris, Brian W. Lee, Ian Waudby-Smith, Philip Amortila, Nika Haghtalab, Michael I. Jordan
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GT
arXiv:2607.26358v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning is widely used in language model training to improve model performance on a target task while limiting drift from a reference policy. A standard way to balance this trade-off is via a KL-regularized RL objective,...
19. ClockRoPE: Random Fourier Rotations for Temporal Routine Modeling ​
Author: Yiwen Chen, Joshua Ainslie, Krzysztof Choromanski, Xiang Gao, Su-Lin Wu, Yiping Yuan, Qian Sun
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26369v1 Announce Type: new Abstract: Rotary Position Embedding (RoPE) has been widely adopted in transformer-based large language models. However, its log-linear frequency schedule, originally designed to produce long-term attention decay, limits its adoption in domains with more complex ...
20. Q-Steer: Action-Value Guidance for Molecular Policy Optimization ​
Author: Xinyu Wang, Jinbo Bi, Minghu Song
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM
arXiv:2607.26391v1 Announce Type: new Abstract: Oracle-limited molecular optimization gives reward only after a complete molecule is generated, while each rollout requires many local next-token decisions. This delayed-feedback interface makes molecular policy optimization myopic: an optimizer can le...
21. Flow Map Learning via Nongradient Vector Flow ​
Author: Mark Goldstein, Anshuk Uppal, Raghav Singhal, Aahlad Puli, Rajesh Ranganath
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26398v1 Announce Type: new Abstract: Diffusion and flow-based models benefit from simple regression losses, but inference incurs significant overhead because sampling requires integration. Consistency models address this by directly learning the flow maps along the ODE trajectory, opening...
22. Examining the Efficacy of Graph Neural Network Message-Passing in Regression Contexts ​
Author: Keith G. Mills, Aedan J. DeFrates, Joong Ho Kim
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26404v1 Announce Type: new Abstract: Graph Neural Networks (GNN) facilitate effective prediction on graph data such as molecules, media networks and neural network blueprints. GNNs facilitate prediction through message passing techniques which define how information flows from a node to i...
23. SCOUT: Per-Context Reset Curricula for Sparse-Reward Reinforcement Learning ​
Author: Siddharth Aphale, Ayushman Singh
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26417v1 Announce Type: new Abstract: Sparse-reward reinforcement learning often fails because rollouts from the unassisted evaluation start rarely reach later task stages. Reset curricula address this by starting some training rollouts from easier intermediate states, called scaffolds. Su...
24. Existence-Field Diffusion Model for Spatial Point Processes with Variable Cardinality ​
Author: Xiaoyin Pan, Christian R. Shelton, Rakshith Mahishi, Chengkuan Hong
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.26428v1 Announce Type: new Abstract: We study generative modeling of spatial point processes (SPP), where both the number of points and their spatial configuration are governed by a joint distribution. While diffusion models have achieved strong performance in modeling complex distributio...
25. DHRCL:Training Code LLMs with Dense Hierarchical Rewards and Curriculum Learning ​
Author: Shuhang Wang, Ziming Li, Hui Cheng
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26457v1 Announce Type: new Abstract: Reinforcement learning is a natural post-training paradigm for code-oriented large language models because generated programs can be evaluated through parsing, execution, unit tests, and structural analysis.However, existing methods often rely on spars...
26. Neural Architecture Search for Traffic Prediction: A Survey of Methods, Challenges, and Future Directions ​
Author: Truong Giang Vu, Li Yang, Richard W. Pazzi
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2607.26467v1 Announce Type: new Abstract: Traffic prediction is a core task in intelligent transportation systems, supporting applications such as adaptive signal control, route guidance, and ride-hailing dispatch. Deep learning models, including graph convolutional networks, recurrent network...
27. Learning Dynamic User Personas from Implicit Interaction Streams via Iterative Refinement ​
Author: Haifeng Wu
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.26473v1 Announce Type: new Abstract: Personalizing large language models (LLMs) to individual users is essential for improving user experience, yet existing approaches typically rely on explicit preference supervision such as pairwise comparisons or demographic attributes, limiting their ...
28. Conformal Changepoint Localization and Root Cause Analysis with Corrupted Observations ​
Author: Seunghun Yu, Meiyi Zhu, Petar Popovski, Joonhyuk Kang, Osvaldo Simeone
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2607.26481v1 Announce Type: new Abstract: Detecting when the statistical behavior of an engineered system changes, and identifying which component is responsible, are core problems in the monitoring of telecommunication networks, robotic platforms, security infrastructure, and multi-agent syst...
29. From Conceptual Hydrologic Models to Conceptually Interpretable Neural Networks: A Snow-Water Mass-Conserving-Perceptron Framework for Discovering Catchment-Scale Precipitation-Storage-Runoff Representations ​
Author: Yuan-Heng Wang, Hoshin V. Gupta
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26492v1 Announce Type: new Abstract: The Mass-Conserving Perceptron (MCP) establishes a modeling paradigm in which conceptual hydrologic models can be reformulated as physically constrained, conceptually interpretable neural networks. Here, we develop a snow-water MCP network framework an...
30. From Interface to Inference: Eliciting Any-Order Inference from Any-Order Models ​
Author: Seunggeun Kim, Jaeyeon Kim, Taekyun Lee, Yuyuan Chen, Yilun Du, Sham Kakade, Sitan Chen
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26504v1 Announce Type: new Abstract: Many discrete reasoning tasks, such as code generation, are inherently non-causal: programmers move between high-level structure and local details, a process we call any-order inference. For autoregressive language models, which lack a native any-order...
31. Collaborative Weighting with Pessimistic Critic for Mitigating Overestimation in Off-Policy Reinforcement Learning ​
Author: Gong Gao, Xiao Lai, Ziqi Xie, Guojie Chen, Xianhui Liu, Weidong Zhao
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26509v1 Announce Type: new Abstract: Deep off-policy reinforcement learning algorithms for continuous control typically rely on neural value function approximation to guide policy improvement. However, temporal-difference (TD) learning introduces noisy targets, resulting in non-stationary...
32. HiFloat4 Format for End-To-End Reinforcement Learning Post-Training of Large Language Models ​
Author: Hei Yi Mak, Shadan Golestan, Hoang Le, Mehran Taghian Jazi, Yunke Peng, Yaoyuan Wang, Yao Wang, Junsong Wang, Tianchi Hu, Fengchen He, Guipeng Hu, Tanzila Rahman, Anandharaju Durai Raju
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26515v1 Announce Type: new Abstract: We present, to our knowledge, the first end-to-end FP4 RL post-training, in which both the rollout and training policies, including their forward and backward passes, operate at 4-bit precision. A systematic study reveals that the dominant source of de...
33. From Unsupervised Subgroups to Hypothetical State-Intervention Policies: An Evaluation of Selected Subgrouping Methods in Observational Health Data ​
Author: Vasundhara Acharya, Bulent Yener
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26521v1 Announce Type: new Abstract: Conventional subgroup analyses can yield unstable and difficult-to-interpret conclusions, especially in observational biomedical data where each individual is observed under only one exposure state, true individual treatment effects are unavailable, an...
34. The Art of Not Forgetting A Local Learning Architecture for Continual Learning ​
Author: Ashmith Atmuri, Yashaswini Rao Bhogarajula
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26523v1 Announce Type: new Abstract: We introduce CMP (Cognitive Memory Primitive), a continual-learning architecture that repre?sents inputs as sparse relational codes, stores them in a two-tier competitive memory, and learns through local updates without end-to-end backpropagation throu...
35. AgentGFM: A Graph Foundation Model with Node-Agent Information-Flow Control ​
Author: Jingbo Cui, Jitao Zhao, Di Jin, Dongxiao He
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26533v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) aim to learn transferable knowledge from multi-domain graphs and adapt to unseen scenarios. As a fundamental source of relational semantics in graphs, the transferability of topological patterns has long been central to G...
36. From Tokens to Watt-hours: Analytical Energy Estimation for LLM Inference on Modern GPUs ​
Author: Tina Vartziotis, Rodopi Kosteli, Elli Vartziotis, George Dasoulas, Michael Keckeisen, Konstantinos Skianis, Sotirios Kotsopoulos, Francesca Dominici
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.SE
arXiv:2607.26571v1 Announce Type: new Abstract: The operational energy consumption of large language model (LLM) inference is becoming an increasingly important component of the environmental footprint of deployed AI systems. However, direct measurement of inference energy often requires hardware te...
37. Simultaneous Coverage and Efficiency Guarantee in Online Conformal Prediction ​
Author: Rahul Vaze
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.DS
arXiv:2607.26577v1 Announce Type: new Abstract: Adaptive conformal inference (ACI) of Gibbs and Cand{`e}s and its variants are the standard approach to online conformal prediction under distribution shift, but they suffer from three fundamental limitations. First, their guarantees control only the ...
38. Benchmarking ConvLSTM for One-Day-Ahead IMDAA Rainfall-Field Prediction across Four Indian Cities ​
Author: Tanmay Ghosh, Shaurabh Anand, Rakesh Gomaji Nannewar, Nithin Nagaraj
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26581v1 Announce Type: new Abstract: Convolutional long short-term memory networks (ConvLSTMs) are widely used for precipitation forecasting, but most evidence for their performance comes from dense, high-frequency radar sequences. This study tests whether convolutional recurrence improve...
39. Uncertainty-Guided LLM Semantic Augmentation for Heterogeneous Treatment Effect Estimation ​
Author: Jialu Xu, Mengkun Liang, Guannan Liu, Xiaojie Mao, Junjie Wu
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26599v1 Announce Type: new Abstract: Estimating heterogeneous treatment effects is central to targeted interventions, such as personalized promotions and precision medicine. We focus on the conditional average treatment effect (CATE), a standard estimand for characterizing such heterogene...
40. FedWeave: Rethinking the Unit of Specialization in Heterogeneous Federated MoE-LoRA ​
Author: Donghang Duan, Xu Zheng, Lizong Zhang, Chong Mu, Meng Han
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.26618v1 Announce Type: new Abstract: Federated PEFT enables LLMs to collaboratively adapt to decentralized private data without sharing raw examples. However, task heterogeneity across clients can cause cross-task interference and gradient conflicts during aggregation. Federated MoE-LoRA ...
41. Enhancing Automated Machine Learning via Homogeneous Train-Test Splitting Methods ​
Author: Yearn Tan Yin Tze, Charles Grellois
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26625v1 Announce Type: new Abstract: Accurate model evaluation in machine learning depends critically on how datasets are split into training and testing subsets. Standard random splitting assumes that both partitions share the same underlying distribution, an assumption often violated in...
42. Understanding Context Sampling in TabPFN on Small Tabular Datasets ​
Author: Mohammed Abdullah
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26628v1 Announce Type: new Abstract: TabPFN performs classification through in-context learning: it conditions on a set of labeled training rows (the context, or prototypes) and predicts test labels without gradient updates. On small tabular datasets, practitioners must still choose the c...
43. RAG-HAR+: Towards Cost-Efficient LLM-Based Human Activity Recognition for Edge Deployment ​
Author: Hansi Karunarathna, Nirhoshan Sivaroopan, Chamara Madarasingha, Anura Jayasumana, Kanchana Thilakarathna
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2607.26631v1 Announce Type: new Abstract: Human Activity Recognition (HAR) from wearable sensors supports applications in healthcare, rehabilitation, fitness tracking, and smart environments. Yet, existing deep learning approaches require dataset-specific training, large labeled corpora, and r...
44. AIGen: Automating AI Bill of Materials Generation Through Hybrid MLOps Integration ​
Author: Federica Pepe, Daniele Bifolco, Costantino Martignetti, Aureliano D'Amici, Fabiano Izzo, Damian A. Tamburri, Massimiliano Di Penta
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26652v1 Announce Type: new Abstract: The responsible development and deployment of artificial intelligence (AI) systems requires rigorous documentation of their constituent artifacts, e.g., datasets, model weights, training pipelines, and runtime dependencies. Although the Software Packag...
45. Efficient Heteroscedastic Bayesian Optimization for Risk-Aware AutoRL ​
Author: Mingxuan Che, Tsung-Yuan Tseng, Theresa Eimer, Marius Lindauer, Alexander von Rohr
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26680v1 Announce Type: new Abstract: Reinforcement learning (RL) has shown remarkable success across a wide range of complex tasks. However, RL outcomes can be highly stochastic, and both expected performance and variability often depend on hyperparameter (HP) configurations. We propose e...
46. Universality and Approximation Rates of Graph Neural Networks with Random Features ​
Author: Lukas Gonon, Thilo Meyer-Brandis, Niklas Weber
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.26699v1 Announce Type: new Abstract: We investigate message-passing graph neural networks with random node features. Random node features are known to enhance the expressiveness of graph neural networks (GNNs) both theoretically and empirically. Here, we establish a novel universality res...
47. Mixture-of-experts for handwriting trajectory reconstruction from IMU sensors ​
Author: Florent Imbert, Eric Anquetil, Yann Soullard, Romain Tavenard
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26708v1 Announce Type: new Abstract: The use of digital pens for online handwriting trajectory reconstruction is a prevalent method for human-computer interaction. In this study, we focus on a digital pen equipped with sensors where we aim at reconstructing the online handwriting trajecto...
48. PowerAtlas: Towards Electricity-Computing Co-Scheduling for Power Systems ​
Author: Kaiwen Jiang, Siya Xu, Ziyue Zhu, Chao Yang, Anh Tuan Luu, Haoran Luo
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26710v1 Announce Type: new Abstract: The rapid growth of AI workloads is turning data centers into large-scale, volatile, yet spatiotemporally flexible grid loads, creating an urgent need for coordinated electricity-computing scheduling. Under stringent grid constraints, schedules from ge...
49. Domain adaptation for handwriting trajectory reconstruction from IMU sensors ​
Author: Florent Imbert, Romain Tavenard, Yann Soullard, Eric Anquetil
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26736v1 Announce Type: new Abstract: Digital pens are commonly used to write on digital devices, providing the handwriting trace and enhancing human-computer interation. This study focuses on a digital pen equipped with kinematic sensors, allowing users to write on any surface while simul...
50. CalTwin: Towards Calibrated, Shift-Robust Medical World Models via Fisher-Information Regularisation ​
Author: Behraj Khan, Shabir Ahmad, Syed Ahmad Chan Bukhari, Tahir Qasim Syed
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26752v1 Announce Type: new Abstract: Medical world models aim to learn a latent state of patient or organ physiology and a transition function that forecasts how that state evolves under interventions, supporting downstream tasks from imaging-based diagnosis to digital-twin treatment plan...
51. Journey Operators for Structured Multi-Axis Composition ​
Author: Mahesh Godavarti
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26775v1 Announce Type: new Abstract: Many kinds of data have structure along one or more axes: words in a sentence, pixels in an image, nodes in a tree, frames in audio, or cells in a 3D volume. Along one axis, order matters: "the dog bit the man" is different from "the man bit the dog." ...
52. SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ​
Author: Zhiyuan Yao, Yuxin Chen, Zhengxi Lu, Zishan Xu, Yueqing Sun, Yifu Guo, Yuquan Lu, Zhengzhou Cai, Kangning Zhang, Zhuowen Han, Zi-Han Wang, Ziang Ye, Qi Gu, Xunliang Cai, Weiwen Liu, Yongliang Shen
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26784v1 Announce Type: new Abstract: Large language model agents often encounter related yet distinct tasks that share reusable solution patterns. Yet standard agentic reinforcement learning treats tasks as independent episodes, while existing approaches to skill learning either focus on ...
53. FedTopo: Relation-Level Topology Sharing for Model-Heterogeneous Federated Learning ​
Author: Zhaoyang Ma, Zhihao Wu, Xin Gao, Lipo Wang, Youfang Lin, Jing Wang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26801v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative learning over decentralized data silos without centralizing raw data. However, heterogeneous local architectures often induce non-aligned representation spaces, making it difficult to transfer global knowle...
54. Forecasting Trajectory-Level Safety Risks in Black-Box Multi-Turn Interactions ​
Author: Shi Lin, Peng Qian, Dinghao Liu, Renjie Sun, Sifan Wu, Dezhang Kong, Chenpei Wang, Xun Wang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2607.26820v1 Announce Type: new Abstract: As large language models (LLMs) evolve from standalone assistants into autonomous agents, ensuring their safety requires shifting beyond pointwise risk assessment to understand how risks emerge and unfold over long-horizon trajectories. In multi-turn i...
55. Budget-Aware LLM Discovery via Cost-Calibrated Frontier Utility ​
Author: Yansen Zhang, Yilu Liu, Tianyu Liu, Jiamin Chen, Xiaokun Zhang, Kai Xie, Xue Liu, Chen Ma, Yiyan Qi
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26828v2 Announce Type: new Abstract: Large language models increasingly support scientific and algorithmic discovery through inference-time search over evaluated candidates. Existing adaptive discovery controllers assign credit based only on score progress, even though prompt length, retr...
56. Kairos: Numerically Robust News Recommendation under Item Cold-Start via Cholesky-based LinUCB ​
Author: Finn Hertsch
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2607.26832v2 Announce Type: new Abstract: Algorithmic news personalization in regional markets often fails because modern deep learning models require massive interaction data while real-world news has a short Time-to-Live (TTL < 48 h) and shallow article pools. This structural item cold-start...
57. Tight Generalization Bound for AdaBoost ​
Author: Mikael M{\o}ller H{\o}gsgaard
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26838v1 Announce Type: new Abstract: In this paper we show that the generalization error of AdaBoost is $\Theta\big(\tfrac{d\ln(n\gamma^{2}/d)}{n\gamma^2}+\tfrac{\ln(1/\delta)}{n}\big)$, where $\gamma$ is the advantage guaranteed by the weak learner, $d$ is the VC-dimension of the class c...
58. Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models ​
Author: Hua-Dong Xiong, Xinyuan Yan, Ji-An Li, Jingming Xue, Marcelo G. Mattar, Robert C. Wilson
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26845v1 Announce Type: new Abstract: Inference-time thinking improves the performance of large language models, but aggregate outcomes do not reveal whether models use available evidence more effectively or seek information that could improve future decisions. We distinguish these respons...
59. TREA-Net: A Transferable Residual Epidemiological Adaptation Network for Dengue Incidence Forecasting ​
Author: Inesh Shukla, Madhurima Panja, Tanujit Chakraborty, Chittaranjan Hens
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2607.26854v1 Announce Type: new Abstract: Accurate multi-week dengue forecasting supports timely vector-control interventions, outbreak preparedness, and healthcare resource allocation. However, newly established surveillance systems often lack the historical data needed to train reliable neur...
60. Amortized Moment Matching for Visual Generation ​
Author: Wenze Liu, Xintao Wang, Pengfei Wan, Xiangyu Yue
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26860v1 Announce Type: new Abstract: We propose amortized moment matching, utilizing neural networks to learn data moments as distributional training signals. By casting diffusion denoisers through polynomial projections, we establish a general framework for moment amortization, revealing...
61. ReCo: Reweighting GRPO Against Distributional Concentration ​
Author: Junoh Park, Junseo Hwang, Wonguk Cho, Taesup Kim
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26862v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a standard reinforcement learning method for post-training language models. Recent work shows that GRPO can reduce the base model's reasoning capacity and underperform it in Pass@k when k is large, i...
62. Actions Have Consequences: Detecting Outcome Performativity using Intervention Testing ​
Author: Brandon Gower-Winter, Georg Krempl
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.26908v1 Announce Type: new Abstract: In many domains such as Palliative Care, Credit Assignment and Recommender Systems, predictions may causally influence the outcomes they predict. This phenomena is known as Outcome Performativity. This paper formalises an approach for detecting Outcome...
63. Two Calls Beat Five Agents: Evaluating Multi-Agent Pipelines Against Self-Refinement for Local Language Models ​
Author: Ashish Prajapati, Om Mohite
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26922v1 Announce Type: new Abstract: Multi-agent LLM pipeline systems break down the task among multiple roles for better reasoning, but are benchmarked mainly with large-scale commercial models. In this study, we investigate Parishad, a structured multi-agent system involving five roles,...
64. Temporally Centered SIGReg Improves Multi-Task LeWorldModel Learning: From Analysis to Method ​
Author: Chang Liu, Fei Suo, Yanzhou Jin, Yusuke Iwasawa, Yutaka Matsuo, Yaonan Zhu
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2607.26924v1 Announce Type: new Abstract: Recent work on LeWorldModel (LeWM) has shown that the Sketched Isotropic Gaussian Regularizer (SIGReg) enables stable end-to-end world-model learning from pixels by regularizing the latent marginal distribution toward an isotropic Gaussian, thereby pre...
65. Surrogate assisted diversity estimation in neural ensemble search ​
Author: Alexandr Udeneev, Petr Babkin, Oleg Bakhteev
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26940v1 Announce Type: new Abstract: Ensembles are a standard way to improve the performance and robustness of deep neural networks, but their effectiveness crucially depends on both the quality and the diversity of individual models. Most neural architecture search (NAS) methods are comp...
66. Foundation Models for Face Presentation Attack Detection: A Unified Linear-Probing Benchmark ​
Author: Peter Lorenz, Anjith George, S'ebastien Marcel
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26993v2 Announce Type: new Abstract: Face presentation attack detection (PAD) remains challenging under cross-dataset evaluation, where domain shift degrades models trained on a single dataset. The scarcity of large-scale labeled data motivates adapting pretrained vision models rather tha...
67. What Can Latent World Models Know? Physical Parameter Identifiability in Multimodal Predictive Representations ​
Author: Kaizhen Tan, Xin Xu, Siru Tao, Hanzhe Hong, Yang Feng, Heqing Du
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2607.27017v2 Announce Type: new Abstract: A central premise of latent world models is that predicting the future forces a representation to internalize the physics of its environment. Which physical quantities does a trained latent actually contain, and what decides this? We answer with contro...
68. BayesAME: Bayesian Active Model Evaluation ​
Author: Paula Cordero Encinar, Taylan Cemgil, Arnaud Doucet, Virginia Aglietti, Silvia Chiappa
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.27023v1 Announce Type: new Abstract: Evaluating large generative models across benchmarks is time-consuming and computationally expensive. This drives the need for methods that can estimate full benchmark performance by evaluating models on only a subset of items, known as a coreset. Curr...
69. TreeCCA: Canonical Correlation Analysis via Gradient-Boosted Trees ​
Author: James Chapman
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27027v1 Announce Type: new Abstract: Gradient-boosted trees dominate tabular machine learning, yet canonical correlation analysis has always relied on linear or neural encoders. We propose \textbf{TreeCCA}, the first method to train gradient-boosted tree ensembles end-to-end as CCA encode...
70. Lottery Tickets Are Not Deployment Tickets ​
Author: Bum Jun Kim
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.27031v1 Announce Type: new Abstract: Reports on how sparsification, compression, and lottery tickets change model behavior have been mixed in the prior literature, with beneficial effects observed in some studies and adverse effects in others. Moreover, prior work has not considered actua...
71. CoCaRS: Correlation Calibration-Based Redundancy Suppression for Heterogeneous Knowledge Distillation ​
Author: Fengming Yu, Haiwei Pan, Kejia Zhang, Chunling Chen, Jian Guan, Baoying Ma
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.27054v1 Announce Type: new Abstract: Knowledge distillation (KD) enables a compact student model to learn from a powerful teacher and has become an effective paradigm for model compression. The emergence of diverse model architectures has extended KD from homogeneous to heterogeneous sett...
72. Parameter-Free Dynamic Regret for Online Convex Optimization under Heavy-Tailed Noise ​
Author: Vaneet Aggarwal
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC
arXiv:2607.27073v1 Announce Type: new Abstract: We study online convex optimization (OCO) in non-stationary environments under heavy-tailed noise, where the stochastic gradient oracle admits only a finite $p$-th central moment for some $p \in (1, 2]$. While static regret is well-understood, achievin...
73. Single-Beat Cuffless Blood Pressure Estimation Using Ear-PPG and ECG with a Lightweight Hybrid Learning Framework ​
Author: Kindeep K. Dhatt, Tengyue Wu, Hanbang Hua, Yayun Du
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SP, eess.SY
arXiv:2607.27076v1 Announce Type: new Abstract: Continuous cuffless blood pressure (BP) monitoring remains challenging due to motion artifacts, physiological variability, and the limited robustness of conventional pulse transit time (PTT) models under dynamic conditions. Many prior approaches rely o...
74. Equilibrium Training of Energy-Based Models with Parallel Trajectory Tempering ​
Author: Nicolas B'ereux, Aur'elien Decelle, Cyril Furtlehner, Beatriz Seoane
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, cond-mat.stat-mech
arXiv:2607.27077v1 Announce Type: new Abstract: Energy-Based Models (EBMs) provide an interpretable framework for generative modeling of scientific data, but poor Markov Chain Monte Carlo mixing often limits their reliability. We introduce a training algorithm based on Parallel Trajectory Tempering ...
75. Scores Are Not Decisions: Cost-Aware Stopping for Tool Acquisition in LLM Agents ​
Author: Yicheng Feng, Yan Zhang, Yan Cheng, Wei Qi
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.27083v1 Announce Type: new Abstract: As LLM agents increasingly depend on diverse external services such as search engines, databases, and connectors, agent harnesses face a fundamental tool-selection challenge: acquiring too few tools leaves the task under-informed, while too many adds c...
76. Sky sphere representation in language models ​
Author: Aleksandr Berdnikov, Yevgeny Liokumovich
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27092v1 Announce Type: new Abstract: We analyze whether language models of size ~100B have a representation of the night sky map that is decodable from their residual stream. We find that most of the considered open-source models do have such a representation, and it often even surfaces t...
77. Hierarchical Spatio-Temporal Transformer for Coherent Emergency Department Forecasting ​
Author: Filipa Lino, B'arbara Tavares, Carlos Santiago, Cl'audia Soares, Manuel Marques
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27106v1 Announce Type: new Abstract: Emergency Departments (EDs) are critical access points in healthcare systems, yet they face persistent pressure from unpredictable patient demand, seasonal surges, and non-urgent visits. Effective ED planning requires forecasts at multiple decision-mak...
78. Voronoi Histograms for Adaptive Vectorization of Expected Persistence Diagrams ​
Author: Kaifeng Zhang, Kai Ming Ting
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27126v1 Announce Type: new Abstract: Persistence Diagram (PD) is known to capture point cloud topology effectively, but its computation has high time complexity. Expected Persistence Diagram (EPD) has been developed to reduce the time cost by studying the topology of multiple subsets of a...
79. Minimal Markovization via Stable Quotients in Holonomy-Cover Decision Processes ​
Author: Zuyuan Zhang, Yongshan Chen, Mahdi Imani, Tian Lan
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27132v1 Announce Type: new Abstract: An agent acting under partial observability must retain a recursively updateable statistic of history that restores the Markov property, but the smallest such statistic is generally unknown. We characterize this minimal Markov sufficient statistic for ...
80. Cost-Sensitive Conformal Prediction and Human-in-the-Loop Abstention for Imbalanced High-Stakes Decision Support: A Multi-Domain Benchmark ​
Author: Manpreet Singh, Akshatha Srikantha, Shyamal Lakhanpal
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.27143v1 Announce Type: new Abstract: High-stakes decision systems in credit scoring, fraud detection, healthcare, and industrial safety require reliable uncertainty quantification under severe class imbalance and asymmetric error costs. Standard marginal conformal prediction (CP) provides...
81. Skillful forecasting of offshore winds from satellite scatterometer constellations ​
Author: Francesco Pinto, Luca Lanzilao, Paco Lopez Dekker, Angela Meyer
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27152v1 Announce Type: new Abstract: Accurate intraday forecasts of offshore wind are becoming increasingly important for power system operation and the integration of growing shares of offshore wind energy. Operational forecasts rely predominantly on numerical weather prediction (NWP), w...
82. When Do Learned Diffusion Proposals Help Constraint Solving? A Controlled Study on Continuous Algebraic Systems ​
Author: Quang Bui, Sparsh Roy, Akash Gundimeda, Davin Yin
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27169v1 Announce Type: new Abstract: Solving a continuous algebraic constraint system requires two decisions: which values satisfy the constraints, and which structural augmentation renders an unsolvable system solvable. Classical solvers answer the first well and the second only by enume...
83. Inverse Learning of Latent Risk-Neutral Densities from Irregular Option Quotes ​
Author: Lennon J. Shikhman, Michael Galarnyk, Aadi Dash, Nicholas A. Welsh
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP, q-fin.PR, q-fin.ST
arXiv:2607.27188v1 Announce Type: new Abstract: Accurate option prices do not imply accurate recovery of the latent risk-neutral density. We study this distinction with two complementary benchmarks. A controlled benchmark exposes simulator-truth densities for latent evaluation, while a chronological...
84. From Classification to Regression: Using a Fruitfly to Solve Equations ​
Author: Shady E. Ahmed, Panos Stinis
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2607.27196v1 Announce Type: new Abstract: We present a novel approach to regression tasks using classification which is motivated by the mechanism used by fruitflies to sense their environment. Specifically, we formulate a general framework for learning nonlinear input-output relationships by ...
85. Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? ​
Author: Perry Dong, Ron Polonsky, Dorsa Sadigh, Chelsea Fin
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.27203v1 Announce Type: new Abstract: Pre-training followed by fine-tuning has become the dominant recipe for learning performant policies, and in value-based reinforcement learning (RL) this raises a natural question: given a pretrained policy, should the Q-function be pretrained on offli...
86. Simplex Demixing: Disentangling Multiple Light-Flavor Jets at Colliders ​
Author: Gregorio de la Fuente, Jesse Thaler
Published: 7/30/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-ex
arXiv:2607.24921v1 Announce Type: cross Abstract: Providing a practical and hadron-level definition of multiple jet flavors has been a long-standing challenge in collider physics. Previous work has introduced a data-driven, operational definition of quark and gluon jets, but no robust generalization...
87. Archetypes or ability? Clustering for modelling student mathematical competence ​
Author: Benjamin Mawdsley, Tom Quilter, Richard Turner, Sarah Jackson, Paul Edwards
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG
arXiv:2607.26063v1 Announce Type: cross Abstract: Personalised learning systems often assume that mathematical ability is combined of discrete abilities, acquired sequentially and dependent upon first acquiring foundational abilities, and students often report different strengths. In this work, we e...
88. When Kernel Ridge Regression Meets the H\"older-Zygmund Class: Minimax Optimality and Failure of Properness ​
Author: Yuxuan Hou
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2607.26065v1 Announce Type: cross Abstract: We study kernel ridge regression for nonparametric regression over the H"older-Zygmund class. Using an RKHS equivalent to a Sobolev space of smoothness s+d/2, we prove that misspecified KRR attains the minimax L2 rate n^{-2s/(2s+d)}. We also show th...
89. Shape-Based Inductive Bias for Glioma Grading from Tumor Contours ​
Author: Puneet Velidi, Michelle F. Miranda, Farouk Nathoo, Ashery Mbilinyi, C'edric Beaulac
Published: 7/30/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2607.26090v1 Announce Type: cross Abstract: Glioma grading from tumor contours is often treated as a pixel problem even when the signal of interest is shape. We align closed contours with a functional shape-alignment framework, separate global deformation from residual Fourier shape, and organ...
90. Lilith: Backdoor Generalization under Training-Inference Trigger Shift ​
Author: Zhou Feng, Jiahao Chen, Chunyi Zhou, Yuan Su, Tianyu Du, Yuwen Pu, Jianhai Chen, Jinbao Li, Shouling Ji
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.26099v1 Announce Type: cross Abstract: Machine-learning services increasingly rely on public data, third-party providers, and outsourced training, creating opportunities for data-poisoning attacks that implant persistent malicious behavior while preserving benign utility. However, existin...
91. Weight and Height Estimation from a Single Human Image Captured in the Wild ​
Author: Hira Yaseen, Arif Mahmood, Waqas Sultani
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.26104v1 Announce Type: cross Abstract: A person's physical characteristics such as weight and height are important indicators of his physical and mental health, daily life routines and finances. Body Mass Index (BMI) is a well known measure that encodes the characteristics of both the wei...
92. Two2Four: Generative Quadruped Puppeteering from Human Motion ​
Author: Fatemeh Zargarbashi, Zehong Qiu, Dhruv Agrawal, Stelian Coros, Robert W. Sumner, Martin Guay, Jakob Buhmann
Published: 7/30/2026, 4:00:00 AM
Categories: cs.GR, cs.LG
arXiv:2607.26108v1 Announce Type: cross Abstract: Realistic animal motion for virtual production is typically obtained either through motion capture of highly trained performers who accurately mimic animal behavior, or by retargeting ordinary human motion using complex control setups. Both approache...
93. GPT-Red: Automated Red Teaming via Self-Play at Scale ​
Author: Eric Wallace, Christopher A. Choquette-Choo, Nikhil Kandpal, Sam Toyer, Dylan Hunn, Stephanie Lin, Yuxin Wen, Xiangyu Qi, Christopher Wolff, Zizhao Wang, Milad Nasr, Sicheng Zhu, Chuan Guo, Juan Felipe Cer'on Uribe, Kaiwen Wang, Aiden Low, Kai Xiao, Kai Chen
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG
arXiv:2607.26115v1 Announce Type: cross Abstract: We introduce \textbf{GPT-Red}, an automated red-teaming agent that is trained to discover novel prompt injection attacks against frontier LLMs. The goal of this model is to evaluate and improve the robustness of our production systems. To this end, w...
94. Try Again, Don't Look Back: Blind Resampling Outperforms Self-Repair in Small Code Models ​
Author: Yuvraj Verma
Published: 7/30/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2607.26117v1 Announce Type: cross Abstract: Self-repair - returning a failed program to the model together with its test output and asking for a correction - is a standard component of code agents, and is almost always evaluated against a baseline that does not retry at all. We argue that this...
95. When benchmark inferences do not compose: Projectibility in AI evaluation ​
Author: Brett Reynolds
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG
arXiv:2607.26159v1 Announce Type: cross Abstract: An AI benchmark result rarely reaches a consequential claim in one step. Evaluators generalize it to further cases, interpret it as evidence of capability, extrapolate it to new tasks, transport it to another system or site, and combine it with assum...
96. A Picture Says Thousands of Words - Harnessing Dermal Exposure Data from Images through Hybrid Deep Learning for Enhanced Safety Assessment ​
Author: Hua Qian, Manisha Kotha, Tuan Tran, Jennifer Shin, Haining Zheng
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.26170v1 Announce Type: cross Abstract: This study developed a hybrid computer vision method to quantify exposed skin from images for dermal exposure assessment. Using 170 indoor-painting images, Mask R-CNN first identified human subjects and removed background interference; a color-based ...
97. Position: Evaluation Scores Are Perishable Knowledge Claims ​
Author: Sankalp Gilda, Shlok Gilda
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.SE
arXiv:2607.26191v1 Announce Type: cross Abstract: Evaluation methodologies for language models increasingly combine multiple signals, from automated metrics and LLM-as-judge ratings to human assessments and benchmark suite results. When these signals are aggregated via averaging, evaluation confiden...
98. Randomizing the Number of Centers in k-means++ ​
Author: Vaclav Rozhon
Published: 7/30/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, stat.ML
arXiv:2607.26202v1 Announce Type: cross Abstract: The $k$-means++ algorithm is a standard and widely used seeding method for $k$-means clustering, but for a fixed number $k$ of centers its worst-case expected approximation ratio is $\Theta(\log k)$. We consider the same algorithm when an adversary f...
99. Retrospective Orthogonal Design: Response-Surface Reconstruction from Observational Data ​
Author: Lawrence Fulton, Christopher Fulton, Arvind Sharma, Aleksandar Tomic
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, math.DG, stat.AP
arXiv:2607.26219v1 Announce Type: cross Abstract: Regression estimates from observational data can depend on specification under multicollinearity, while sequential sums of squares (SS) depend on term order. We introduce Retrospective Orthogonal Design (ROD), which reconstructs conditional mean surf...
100. Lightweight Image Classification of Raptor Species for Edge Devices: Rare-Species Dataset Expansion via Video Frame Extraction, Knowledge Distillation, and TensorRT Deployment ​
Author: Takeshi Nishikawa
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV
arXiv:2607.26238v1 Announce Type: cross Abstract: We investigate lightweight raptor-species classification for real-time edge deployment in wind-turbine collision mitigation. Using DINOv2-L (304M parameters) as a teacher, we distilled three lightweight students (MobileNetV4, ViT-Small, and Efficient...
101. Learning the Word Problem: Geodesic Lengths and Cryptographic Applications ​
Author: Elisabeth Fink
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, math.GR
arXiv:2607.26241v1 Announce Type: cross Abstract: The Word Problem has been a subject of intensive mathematical study for over a century, initially driving advances in combinatorial group theory and more recently emerging as a foundational hardness assumption in post-quantum cryptography (PQC). Whil...
102. Denoising growth complexity: Data geometry and certified schedules for diffusion sampling ​
Author: Martin J. Wainwright
Published: 7/30/2026, 4:00:00 AM
Categories: math.ST, cs.LG, cs.NA, math.NA, stat.ML, stat.TH
arXiv:2607.26285v1 Announce Type: cross Abstract: Two central challenges in diffusion-based sampling are the theoretical one of understanding their remarkable effectiveness even in high-dimensional settings, and the practical one of designing algorithms with certified performance guarantees. We show...
103. Rethinking Clinical Relevance in Chest X-ray Machine Learning: How Evaluation References Define Performance ​
Author: Panagiotis Fytas, Ian Selby, Clemens Karner, Judith Babar, Simon Baker, Jake Beckford, Timothy J. Sadler, Shahab Shahipasand, Arthikkaa Thavakumar, John Li Chen, Alex Sawer, Michael Roberts, Jonathan Weir-McCall, J. H. F. Rudd, Carola-Bibiane Sch"onlieb, Anna Korhonen, Anna Breger
Published: 7/30/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2607.26333v1 Announce Type: cross Abstract: Chest X-ray (CXR) machine learning relies heavily on automated evaluation using reference standards that aim to approximate clinical judgment. However, commonly used report-derived labels for pathology classification or generic image quality metrics ...
104. Incast-Free MoE Rate-Based Scheduling ​
Author: Evyatar Cohen, Jose Yallouz, Alexander Shpiner, Mark Silberstein, Sylvia Ratnasamy, Isaac Keslassy
Published: 7/30/2026, 4:00:00 AM
Categories: cs.NI, cs.DC, cs.LG
arXiv:2607.26340v1 Announce Type: cross Abstract: Mixture of Experts (MoE) architectures have become key to large language models; however, their typical round-robin (RR) scheduling introduces significant bottlenecks. In this paper, we demonstrate that RR causes a previously-undiscovered exponential...
105. Self-Adaptive Learning and Model Predictive Control for Tracking Unknown Dynamics with No Regret ​
Author: Atharva Navsalkar, Hongyu Zhou, Vasileios Tzoumas
Published: 7/30/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SY, eess.SY
arXiv:2607.26370v1 Announce Type: cross Abstract: We propose a self-adaptive online learning for control method for tracking unknown target dynamics. The target dynamics can exhibit switching behavior, particularly, a mixture of structured, random, and/or adversarial motion. Such challenging target ...
106. Knowledge before Reasoning: EC-Reason-Bench, a Training-Free Diagnostic Benchmark for LLM Enzyme Classification ​
Author: Linyu Li, Zhi Jin, Yichi Zhang, Dongming Jin, Yuanpeng He, Huanyao Zhang, Xuan Zhang, Gadeng Luosang, Nyima Tashi
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, q-bio.QM
arXiv:2607.26397v1 Announce Type: cross Abstract: Enzyme function prediction is a hierarchical, knowledge-intensive form of protein function classification. Existing benchmarks expose an anomaly: general LLMs often get the coarse first level right, yet once asked for a complete EC number their accur...
107. Origins and mitigation of double descent in reduced order modeling ​
Author: Andrei A. Klishin, J. Nathan Kutz, Krithika Manohar
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.DS, physics.data-an
arXiv:2607.26414v1 Announce Type: cross Abstract: Latent low-dimensional structure in datasets of natural and engineered systems enables their sparse sensing, or full-state reconstruction from historical data and very few carefully chosen localized measurements. Depending on the reconstruction algor...
108. Chaos Is a LADDER: Domain Generalization Beyond Invariance via Reweighting ​
Author: Yuhang Jiang, Fengchuan Zhang, Sanguo Zhang, Guojun Zhu
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.ME, stat.TH
arXiv:2607.26458v1 Announce Type: cross Abstract: Domain generalization (DG) aims to learn from multiple source domains and generalize to unseen target domains. Most DG methods pursue invariance: they seek a causal representation whose prediction rule is invariant across domains. This principle is e...
109. Parameterized Fair Resource Allocation under Diversity Constraints ​
Author: Keke Huang, Yik Yu Ng, Laks V. S. Lakshmanan, Xiaokui Xiao
Published: 7/30/2026, 4:00:00 AM
Categories: cs.SI, cs.LG
arXiv:2607.26485v1 Announce Type: cross Abstract: Resource allocation across multiple agent groups arises in many applications including e-commerce recommendation systems, housing assignment, and course allocation, and is commonly formulated as an optimization problem with diversity constraints to e...
110. EvoPINN: Agentic Discovery of Executable Algorithms for Physics-Informed Neural Networks ​
Author: Peng Yin, Kai Li, Yifan Zhang, Jian Cheng
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NE
arXiv:2607.26490v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful paradigm for solving partial differential equations (PDEs), yet their performance heavily relies on the manual, trial-and-error engineering of neural representations, loss formulatio...
111. LLMET: Enabling Cross-Layer Evaluation of Emerging M3D Memories for Energy-Efficient LLM Serving ​
Author: Ming-Yen Lee, Hanchen Yang, Faaiq Waqar, Harsono Simka, Tushar Krishna, Muhammed Ahosan Ul Karim, Shimeng Yu
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AR, cs.AI, cs.ET, cs.LG
arXiv:2607.26491v1 Announce Type: cross Abstract: The energy consumption of Large Language Model (LLM) serving is becoming a major system challenge as deployment scales, driven by hardware power and thermal constraints and rising electricity costs. A key contributor to chip energy dissipation is dat...
112. A Persona-based Rate Action Index ​
Author: Hayden Helm, Andrew Dassori
Published: 7/30/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.LG
arXiv:2607.26545v1 Announce Type: cross Abstract: We propose an index for predicting the U.S.\ Federal Open Market Committee (FOMC) decision to hike/hold/cut the current federal funds target rate based on how a collection of personas responds to current market conditions. To construct the index, we ...
113. Adaptive Gradient-Based Methods for a Broader Class of Optimization Problems under Performative Prediction ​
Author: Hiroki Hamaguchi, Yuya Hikima, Hiroshi Sawada, Akiko Takeda
Published: 7/30/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2607.26562v1 Announce Type: cross Abstract: We study optimization under performative prediction, where deploying a model affects the future data distribution. For this setting, several gradient-based approaches have been proposed. However, they typically assume specific data distributions or l...
114. Recover, Decode, Reguard: Guard-Agnostic Defense Amplification againstEncoded VLM Jailbreaks ​
Author: Haoyu Zhang, Zhuoxi Wang, Shibo Zheng, Zijian Xiao, Xiangchen Guan, Mohammad Zandsalimy, Shanu Sushmita
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.26574v1 Announce Type: cross Abstract: Safety classifiers ("guards") are the dominant black-box defense for vision-language models, yet they judge an input's surface form, not its meaning: a harmful request re-encoded as set theory, formal logic, a rare language, code, or an image of text...
115. Classification of Disease from Lungs X-ray Images using VGG16, VGG19 and ResNet50 Models ​
Author: Nand Lal Yadav, Rajesh Kumar, Satyendra Singh, Sudhakar Singh
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.26580v1 Announce Type: cross Abstract: With the increase in the number of cases related to respiratory diseases, there is an urgent need to detect them early and diagnose them accurately. Convolutional neural networks have given promising results when used for diagnosing diseases using im...
116. Harnessing Large Language Models for Intelligent Resource Allocation in the Internet of Everything ​
Author: Haijun Zhang, Zhuojun Duan, Zijun Wu, Xu Ma, Yuzheng Ren
Published: 7/30/2026, 4:00:00 AM
Categories: cs.NI, cs.LG
arXiv:2607.26602v1 Announce Type: cross Abstract: The rapid development of the Internet of Everything (IoE) is accelerating the adoption of intelligent applications. However, the massive number of connected devices generates diverse and heterogeneous tasks, which pose increasing challenges for dynam...
117. Few-Shot Open-Set Audio Classification via Transductive Prototype Refinement and Class Logit Enhancement ​
Author: Tianyan Deng, Yanxiong Li, Rui Gao, Jiahao Du
Published: 7/30/2026, 4:00:00 AM
Categories: cs.SD, cs.LG
arXiv:2607.26607v1 Announce Type: cross Abstract: Few-shot Open-set audio classification requires classifying query samples from known classes with a few labeled support samples while rejecting query samples from unknown classes. Transductive inference jointly observes the full unlabeled query set t...
118. Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting ​
Author: Hongqiang Lin, Chao Liu, Xiaofan Bai, Xuan Jin, Yuhong Li, Nenggan Zheng, Xipeng Cao
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.26643v1 Announce Type: cross Abstract: Enabling large language model (LLM) agents to accumulate and reuse experience from past interactions remains a central challenge in real-world applications. A promising solution is to treat skills as trainable states and optimize them in the same way...
119. The Sparsity Ceiling: Where Spiking Networks Can and Cannot Trade Activity for Energy ​
Author: Zeyu Wang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.NE, cs.LG
arXiv:2607.26648v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) are promoted as an energy-efficient substrate because sparse, event-driven activity replaces dense multiply-accumulates with cheap accumulates. We argue the energy dividend of sparsity is not a property of SNNs but of t...
120. Constitutional Midtraining: Content Presence Drives Alignment Gains ​
Author: Desiree Cho, Cameron Tice, Bernie Hogan, Hunar Batra, Puria Radmard, Jun Zhao, Nigel Shadbolt
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG
arXiv:2607.26654v2 Announce Type: cross Abstract: Post-training alignment is often shallow, eroding under fine-tuning. It remains untested as to whether constitutional midtraining interventions can produce durable alignment when cleanly isolated from post-training. We build a 394M-token constitution...
121. An Informativeness-based Clustered Federated Learning Method for Reliable Traffic Prediction in Managed Wi-Fi Networks ​
Author: Luca Barbieri, Gianluca Fontanesi, Lorenzo Galati Giordano, Alfonso Fernandez Duran, Thorsten Wild
Published: 7/30/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2607.26682v1 Announce Type: cross Abstract: Centrally-managed Wi-Fi solutions are increasingly leveraging Distributed Artificial Intelligence (AI) to predict key operational statistics of Access Points (APs) and proactively optimize network performance. In this context, Clustered Federated Lea...
122. Early Failure Prediction from Near-Anomaly Detection: A Proactive Approach ​
Author: L'ea Billet (LAAS, INSA Toulouse, ANITI), Louise Trav'e-Massuy`es (LAAS-DISCO, Comue de Toulouse, ANITI), Elodie Chanthery (LAAS), Alexandre Gaffet
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.26704v1 Announce Type: cross Abstract: Anomaly detection methods often have uncertain behavior with respect to samples near the distribution boundary, limiting their ability to anticipate future anomalies. This work introduces the concept of near-anomalies that, while not yet anomalous, l...
123. DREvo: Distilling Recalibrated Historical Experience for Harness Self-Evolution ​
Author: Hanghui Guo, Weijie Shi, Zhangze Chen, Shengxiang Xu, Yishu Wang, Yimei Zhang, Wangze Ni, Jia Zhu, Shimin Di
Published: 7/30/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2607.26722v1 Announce Type: cross Abstract: Harness plays a critical role in large language model agent performance, and building a high-performing harness requires substantial expert effort. Therefore, recent research has increasingly explored harness self-evolution, which iteratively propose...
124. Online Handwriting Trajectory Reconstruction from Kinematic Sensors using Temporal Convolutional Network ​
Author: Wassim Swaileh, Florent Imbert, Yann Soullard, Romain Tavenard, Eric Anquetil
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.26733v1 Announce Type: cross Abstract: Handwriting with digital pens is a common way to facilitate human-computer interaction through the use of Online Handwriting (OH) trajectory reconstruction. In this work, we focus on a digital pen equipped with sensors from which one wants to reconst...
125. An Attention-Based Framework for Alzheimers Disease Classification Using Resting-State fMRI ​
Author: Harshiddhi Pathak, Gowtham Reddy N, Mrinal Acharya, Manjunatha Mahadevappa
Published: 7/30/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.HC, cs.LG, eess.SP
arXiv:2607.26746v1 Announce Type: cross Abstract: Accurate identification of Alzheimers disease (AD) using resting-state functional magnetic resonance imaging (rs-fMRI) remains challenging due to the high dimensionality, noise, and complex inter-regional dependencies inherent in functional brain con...
126. Metis: Memory Foundation Model ​
Author: Zeyu Zhang, Ziliang Guo, Yihang Sun, Xichong Zhang, Xixuan Hao, Zehao Lin, Yang Zhang, Xiaoyan Zhao, Tong Shen, Bo Tang, Zhi-Qin John Xu, Junchi Yan, Haofen Wang, Xu Chen, Feiyu Xiong, Zhiyu Li, Tat-Seng Chua
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.26760v1 Announce Type: cross Abstract: Recent advances in AI agents have increasingly internalized native capabilities into their underlying foundation models, giving rise to multimodal foundation models and large reasoning models. However, agent memory is still primarily implemented thro...
127. Stable and Budget-Feasible Coalition Formation for Clustered Federated Learning: A Hedonic Potential-Game Approach ​
Author: Cengis Hasan
Published: 7/30/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2607.26788v1 Announce Type: cross Abstract: Clustered federated learning benefits from organizing heterogeneous participants into coalitions that train coalition-specific models, but such clustering is sustainable only if participants prefer their assigned coalition and the required transfers ...
128. Crossing-Free Probabilistic K-Line Forecasts Without Retraining ​
Author: Runyao Yu, Yuchen Tao, Yujie Chen, Wentao Wang, Derek W. Bunn
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.CE, cs.LG, q-fin.CP
arXiv:2607.26792v1 Announce Type: cross Abstract: Probabilistic K-line forecasting describes uncertainty in four complementary prices, namely open--high--low--close (OHLC). However, it introduces two consistency problems: quantile crossing and K-line crossing. Quantile crossing occurs when a higher-...
129. BATS: Resource-Efficient Volumetric Segmentation with Boundary-Aware Mixed-Resolution Tokens ​
Author: David Hagerman, Roman Naeem, Fredrik Kahl
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.26829v1 Announce Type: cross Abstract: Many high-performing volumetric segmentation models maintain dense multi-scale feature maps, leading to high activation memory and inference cost. We present BATS (Boundary-Aware Token Selection), a 3D medical image segmentation architecture that con...
130. ToxScreen: Detecting Whether an LLM Has Been Poisoned ​
Author: Anthony Hughes, Nicole Xing, Collin Francel, Andy Kim, Andrew Draganov
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.26849v1 Announce Type: cross Abstract: As large language models (LLMs) are deployed in high-stakes domains, adversaries may poison training data to implant backdoors: hidden triggers that covertly manipulate model behavior at inference time. We ask whether a defender can recover such a tr...
131. No Data Is Not No Risk: Visibility Aware Graph-Based Inference of Business Conduct Risk ​
Author: Tsuyoshi Iwata, Johannes Laurmaa, Ryohei Hisano
Published: 7/30/2026, 4:00:00 AM
Categories: q-fin.RM, cs.LG
arXiv:2607.26859v1 Announce Type: cross Abstract: The monitoring of business conduct risk is hindered by sparse, uneven, and visibility-biased data. Prior studies show that business conduct risk information and media coverage propagate through supply chain, peer, and corporate structure networks, ye...
132. Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents ​
Author: Amirmohammad Farzaneh, Osvaldo Simeone
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.IT, cs.LG, math.IT
arXiv:2607.26865v1 Announce Type: cross Abstract: LLM agents following the ReAct paradigm are promising enablers of complex multi-step tasks, including multi-hop question answering, code generation, and control of physical AI systems. Yet, when deployed at the edge, they must tightly manage their re...
133. Conformalized Rate-Adaptive Sensing ​
Author: Jiawei Yang, Yao Zhang
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP, stat.ME
arXiv:2607.26887v1 Announce Type: cross Abstract: Many high-resolution imaging systems face the same fundamental question: when have enough measurements been collected to reconstruct an image accurately? We develop Conformalized Rate-Adaptive Sensing (CoRAS), a method that adaptively chooses an acqu...
134. Same Evidence, Different Target: Decoding How Diagnostic Evidence Bears on Causal Questions from Language-Model States ​
Author: Weiyi Kong, Zhuoran Li
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.26929v1 Announce Type: cross Abstract: The same diagnostic result can support or challenge one causal claim yet fail to address another when the claims concern different populations, outcomes, estimands, pathways, or identifying assumptions. When the evidence and target vary together, a c...
135. Defending Against Backdoor Attacks via Alignment Checking in Model-Contrastive Federated Learning ​
Author: Hongliang Zhang, Zhongyuan Yu, Guijuan Wang, Tianqing He, Wenshuo Ma, Xiaosong Zhang, Jiguo Yu
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.26933v1 Announce Type: cross Abstract: Federated Learning (FL) is vulnerable to backdoor attacks because of its distributed nature in edge computing scenarios. Existing defense methods show limited efficacy as they overlook the deviations among benign local updates caused by statistical h...
136. Breaking the Curse with BAND: Nonparametric Distribution Estimation in High Dimensions ​
Author: Shuo-Chieh Huang, Chien-Ming Chi, Jau-er Chen
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2607.26955v1 Announce Type: cross Abstract: Minimax-optimal rates for multivariate distribution estimation are known to suffer from the curse of dimensionality. We propose a sparse Bayesian network approach in which each conditional probability is estimated using sparsity-aware conditional mea...
137. Feature Bagging Provides Stability ​
Author: Yuheng Ma, Qiang Sun
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2607.26964v1 Announce Type: cross Abstract: We study feature bagging through the lens of algorithmic stability. Feature bagging is an ensemble strategy that aggregates base learners trained on randomly subsampled feature subsets, possibly in a data-dependent manner. We introduce feature instab...
138. Using large language models to probe the limits of atom-centered structural descriptors ​
Author: Michelangelo Domina, Michele Ceriotti
Published: 7/30/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG
arXiv:2607.26984v1 Announce Type: cross Abstract: Mapping an atomic structure to a compact set of geometric descriptors is an essential step in any machine-learning application to atomic-scale modeling. A powerful and widely-used approach can be understood as a discretization of the histogram of pai...
139. SymmGrid: Super-Scaling On-Robot Learning with Parallelized Symmetries and Egocentric-Exocentric Visual Perception ​
Author: Gabe Everett, Brice Gunter, Ryan Vander Stelt, Cleiver Ruiz-Martinez, Blake Hull, Juan Rojas
Published: 7/30/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2607.26985v1 Announce Type: cross Abstract: Deep reinforcement policy learning directly in physical robots (on-robot learning) remains bottlenecked by slow wall-clock training times. We present SymmGrid, a trajectory level augmentation framework inspired by parallelized symmetries that super-s...
140. A Compositional Theory of Causally Masked Transformers ​
Author: Franz Nowak, Ryan Cotterell, Reda Boumasmoud
Published: 7/30/2026, 4:00:00 AM
Categories: cs.FL, cs.LG
arXiv:2607.26988v1 Announce Type: cross Abstract: What types of decision problems can a causally masked, finite-precision transformer solve for inputs of arbitrary length? Existing answers often rely on idealized arithmetic, but under finite precision, rounding and evaluation order can change what i...
141. AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents ​
Author: Ruoyu Wang, Heng Zhao, Renjie Wu, Mengnan Zhao, Zhixuan Chu, Wanyu Lin, Tianhang Zheng
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG
arXiv:2607.26998v1 Announce Type: cross Abstract: Large language model (LLM) agents automate penetration testing through an observation-action loop, selecting actions based on observations returned by tools. This dependence allows defenders to inject deceptive observations that can mislead the agent...
142. On the robustness of noisy solutions in non-convex neural networks ​
Author: Enrico M. Malatesta, Alessandra Passalacqua, Riccardo Zecchina
Published: 7/30/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cs.LG, math.PR
arXiv:2607.27000v1 Announce Type: cross Abstract: Optimization in non-convex neural network models is strongly influenced by the geometry of the solution space: sparse, isolated, point-like clusters are typically algorithmically inaccessible, whereas wide and flat regions can be found efficiently de...
143. HoF-Bench: Rediscovering Real AI-Discovered CVEs Without Frontier Models ​
Author: Petr Simecek, Elnaz Babayeva, Jiri Balhar, Michal Bida, Michal Buran, Vaclav Cadek, Luigino Camastra, Tomas Dulka, Michal Janocko, Tomas Klohna, Pavel Kohout, Ondrej Kokes, Adam Krivka, Jakub Kubik, Patrik Mada, Igor Morgenstern, Marek Pavelka, Joshua Rogers, Petr Stastny, Jan Tattermusch, Dmitrijs Trizna, Martin Votruba, Guido Vranken, Jakub Zikl, Evelina Gabasova, Stanislav Fort
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.27030v1 Announce Type: cross Abstract: LLM-based analyzers have begun finding real vulnerabilities in mature open-source projects: AISLE's analyzer is credited with more than 280 CVEs across 78 projects, including OpenSSL, curl, and GnuTLS. We introduce HoF-Bench (named after AISLE's publ...
144. Mitigating Compounding Error via Video Representation Regularization ​
Author: Taiye Chen, Qi Zhang, Yisen Wang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.27036v1 Announce Type: cross Abstract: Video diffusion-based world models enable long autoregressive video generation for robotics, autonomous driving and simulation tasks, yet sliding-window autoregressive inference suffers from severe error accumulation that degrades frame quality over ...
145. GPTQ-2D: Cubic-Time Two-Sided Adaptive Rounding ​
Author: Jiale Chen, Torsten Hoefler, Dan Alistarh
Published: 7/30/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2607.27042v1 Announce Type: cross Abstract: Adaptive rounding methods such as GPTQ, or equivalently Babai's nearest plane algorithm, round a real matrix to integers under a quadratic metric. They process the entries in a fixed order, one at a time, propagating each rounding error to the entrie...
146. PIKS: Universal Physics-Informed Kernel Methods ​
Author: Joachim Bona-Pellissier, Giacomo Meanti, Matteo Santacesaria, Lorenzo Rosasco
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.27062v1 Announce Type: cross Abstract: Physics-informed machine learning incorporates physical principles --often expressed via differential operators-- into data-driven models. While physics-informed neural networks (PINNs) dominate empirical applications, the complexity of neural networ...
147. Field Codes for Distributed Coupling Samplers and Certified Empirical Transport ​
Author: Hung Mai, Hai Nguyen, Luong Doan, Ngoc Vu, Khanh Nguyen, Nhung Duong, Tuan Do
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CC, cs.IT, cs.LG, math.IT, math.OC
arXiv:2607.27078v2 Announce Type: cross Abstract: In this paper, we formulate three communication tasks for empirical optimal transport: distributed coupling sampling, cost-evaluable coupling output, and scalar value-certified sampling. Our main result is a field-code compiler: any communicated tran...
148. On-Policy Distillation for LLM Safety: A Routing Approach to Template-Robust Realignment ​
Author: Yongjian Guo, Wanlun Ma, Lingyu Shen, Xi Xiao, Sheng Wen
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CR, cs.LG
arXiv:2607.27081v1 Announce Type: cross Abstract: Fine-tuning is the dominant paradigm for specializing large language models (LLMs), yet it exposes a critical vulnerability: malicious data providers can embed harmful behaviors into downstream corpora, creating models that retain professional skills...
149. InferScale: GPU-Native KV Injection for Personalized LLM Serving ​
Author: Peter Li, Prashant Pandey
Published: 7/30/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2607.27090v1 Announce Type: cross Abstract: Large language models are increasingly deployed with persistent personalized context, such as accumulated memory profiles or long conversation histories, that is shared across a user's many requests. Production memory systems (e.g., Mem0, MemGPT, and...
150. Detecting seizure onset and offset times using human intelligence: A critical-transitions-based approach ​
Author: Andrew Flynn, Cian McCafferty, Klaus Lehnertz, Fran\c{c}ois David, Vincenzo Crunelli, Gordon Lightbody, Sebastian Wieczorek
Published: 7/30/2026, 4:00:00 AM
Categories: math.DS, cs.LG
arXiv:2607.27105v1 Announce Type: cross Abstract: Most existing seizure detection algorithms require extensive pre-processing of the data and rely on heuristic or currently unexplainable machine learning approaches. These approaches often struggle with balancing detection sensitivity and specificity...
151. Investigating reservoir computing for branch predictionin pipelined processors using emerging CMOS memristor devices ​
Author: Harvey Samuel George Johnson, Sendy Phang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AR, cs.CE, cs.ET, cs.LG, physics.app-ph
arXiv:2607.27140v1 Announce Type: cross Abstract: This project aimed to develop a novel reservoir compute (RC) implementation framework targeting high-speed operation and integration with CMOS digital logic. With the target workload of branch prediction (BP) for multistage pipelined central pro-cess...
152. MindForge: Teaching Small Language Models Whole-Life-Cycle Software Engineering via Source-Free Program Synthesis ​
Author: Yihao Chen, Shi Chang, Khaled Chawa, Feng Lin, Boyuan Chen, Shaowei Wang, Ahmed E. Hassan
Published: 7/30/2026, 4:00:00 AM
Categories: cs.SE, cs.CL, cs.LG
arXiv:2607.27146v1 Announce Type: cross Abstract: Coding agents have made substantial progress on software engineering tasks that modify existing codebases, including bug fixing and feature implementation. However, constructing a complete program from scratch remains a major challenge: even the fron...
153. Can AI agents conduct open-ended AI research? Early evidence from two case studies ​
Author: Peter Kirgis, Sayash Kapoor, Andrew Schwartz, Stephan Rabanser, David Africa, Konstantinos Voudouris, Viet Nguyen, Toby Pilditch, Magda Dubois, Harry Coppock, Cozmin Ududec, Nitya Nadgir, Matilda Orona, Tilman Bayer, Derrick Chan-Sew, Yue Ling, Abhishek Shetty, Helen Toner, Gillian Hadfield, Seth Lazar, Steve Newman, Shoshannah Tekofsky, Rishi Bommasani, Arvind Narayanan
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG
arXiv:2607.27191v1 Announce Type: cross Abstract: Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carry out open-ended AI research is thin. Current evaluations either test agents on narrow, verifiable tasks, which excludes open-ended r...
154. AI Alignment in Medical Imaging: Unveiling Hidden Biases Through Counterfactual Analysis ​
Author: Haroui Ma, Francesco Quinzan, Theresa Willem, Stefan Bauer
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, eess.IV, stat.ML
arXiv:2504.19621v2 Announce Type: replace Abstract: Machine learning (ML) systems for medical imaging have demonstrated remarkable diagnostic capabilities, but their susceptibility to biases poses significant risks, since biases may negatively impact generalization performance. In this paper, we int...
155. Low-Precision Training of Large Language Models: Methods, Challenges, and Opportunities ​
Author: Zhiwei Hao, Jianyuan Guo, Li Shen, Yong Luo, Han Hu, Guoxia Wang, Dianhai Yu, Yonggang Wen, Dacheng Tao
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2505.01043v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved impressive performance across various domains. However, the substantial hardware resources required for their training present a significant barrier to efficiency and scalability. To mitigate this challeng...
156. Challenges and proposed solutions in modeling multimodal medical data: A systematic review ​
Author: Maryam Farhadizadeh, Maria Weymann, Michael Bla{\ss}, Johann Kraus, Christopher Gundler, Sebastian Walter, Noah Hempen, Hannah Bast, Harald Binder, Nadine Binder
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2505.06945v5 Announce Type: replace Abstract: Multimodal data modeling has emerged as a powerful approach in clinical research, enabling the integration of diverse data types such as imaging, genomics, wearable sensors, and electronic health records. Despite its potential to improve diagnostic...
157. Dense Local Dependencies Induce Attention-Logit Explosion and Training Instability During Long-Sequence Transformer Training ​
Author: Suvadeep Hajra
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2505.15548v2 Announce Type: replace Abstract: Autoregressive transformer language models frequently exhibit training instability when trained on long sequences, particularly under low-precision arithmetic. Although this instability is often accompanied by attention-logit explosion, its underly...
158. Equivariant Eikonal Neural Networks: Grid-Free, Scalable Travel-Time Prediction on Homogeneous Spaces ​
Author: Alejandro Garc'ia-Castellanos, David R. Wessels, Nicky J. van den Berg, Remco Duits, Dani"el M. Pelt, Erik J. Bekkers
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2505.16035v3 Announce Type: replace Abstract: We introduce Equivariant Neural Eikonal Solvers, a novel framework that integrates Equivariant Neural Fields (ENFs) with Neural Eikonal Solvers. Our approach employs a single neural field where a unified shared backbone is conditioned on signal-spe...
159. ReDiSC: A Reparameterized Masked Diffusion Model for Scalable Node Classification with Structured Predictions ​
Author: Yule Li, Yifeng Lu, Zhen Wang, Zhewei Wei, Yaliang Li, Bolin Ding
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2507.14484v2 Announce Type: replace Abstract: In recent years, graph neural networks (GNN) have achieved unprecedented successes in node classification tasks. Although GNNs inherently encode specific inductive biases (e.g., acting as low-pass or high-pass filters), most existing methods implic...
160. The Advantage of Fine-Grained Training ​
Author: Davide Pirovano, Federico Milanesio, Michele Caselle, Piero Fariselli, Matteo Osella
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.05130v2 Announce Type: replace Abstract: In classification problems, models are trained to predict a class label based on the input data features. However, class labels are organized hierarchically in many datasets. While a classification task is often defined at a specific level of this ...
161. Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation ​
Author: Xin Yu, Cong Xie, Xunmei Liu, Tiantian Fan, Lingzhou Xue, Zhi Zhang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.00192v3 Announce Type: replace Abstract: Low-rank adaptation (LoRA) has become a widely used paradigm for parameter-efficient fine-tuning of large language models, yet its representational capacity often lags behind full fine-tuning. Within the context of LoRA, a key open question is how ...
162. On the Rademacher Complexity of Graph Neural Networks: Unifying Expressivity and Geometry ​
Author: Martin Carrasco, Caio F. Deberaldini Netto, Vahan A. Martirosyan, Ehimare Okoyomon, Caterina Graziani
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.10101v4 Announce Type: replace Abstract: Understanding the interplay between generalization, expressivity, and the geometry of the input space is a central challenge in graph learning. The expressivity of Graph Neural Networks (GNNs) is typically characterized through their correspondence...
163. On-Device Inference versus Wireless Streaming: Energy-Efficient Multi-Modal Deep Learning for Wearable Cardiovascular Patches ​
Author: Mustafa Fuad Rifet Ibrahim, Tunc Alkanat, Felix Manthey, Maurice Meijer, Alexander Schlaefer, Peer Stelldinger
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2510.18668v5 Announce Type: replace Abstract: Wearable cardiovascular sensor patches promise continuous, unobtrusive monitoring, but their tight energy, memory, and compute budgets make it unclear whether physiological signals should be analyzed on the device or streamed to the cloud for proce...
164. Revisiting Forest Proximities via Sparse Leaf-Incidence Kernels ​
Author: Adrien Aumon, Guy Wolf, Kevin R. Moon, Jake S. Rhodes
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, cs.PF
arXiv:2601.02735v4 Announce Type: replace Abstract: Decision forests induce supervised similarities through the partition structure of their trees. Yet forest proximity computation is still often treated as a quadratic operation in the number of samples, which limits scalability and restricts broade...
165. Structurally Separated Uncertainty in Supervised Latent Variable Models ​
Author: Tanmoy Mukherjee, Marius Kloft, Pierre Marquis, Zied Bouraoui
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.11219v2 Announce Type: replace Abstract: Predictive uncertainty is commonly decomposed into epistemic and aleatoric components, but standard decompositions often produce strongly correlated estimates because both quantities are derived from the same predictive distribution. We study an al...
166. MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs ​
Author: Ziqiao Shang, Ling-Yue Ge, Zian Xu, Zi-Jian Cheng, Shi-Yu Tian, Zhenyu Huang, Wenbo Fu, Weiming Wu, Yang Chen, Xiangwen Zhang, Yulan Hu, Bin Liu, Lan-Zhe Guo
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.18600v5 Announce Type: replace Abstract: Systematically evaluating Multimodal Large Language Models (MLLMs) is essential for advancing Artificial General Intelligence (AGI). Yet existing benchmarks remain inadequate for rigorously measuring their reasoning capabilities under multi-criteri...
167. Understanding LoRA as Knowledge Memory: An Empirical Analysis ​
Author: Seungju Back, Dongwoo Lee, Naun Kang, Taehee Lee, S. K. Hong, Youngjune Gwon, Sungjin Ahn
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.01097v5 Announce Type: replace Abstract: Continuous knowledge updating for pre-trained large language models (LLMs) is increasingly necessary yet remains challenging. Although inference-time methods like In-Context Learning (ICL) and Retrieval-Augmented Generation (RAG) are popular, they ...
168. Gated Adaptation for Continual Learning in Human Activity Recognition ​
Author: Reza Rahimi Azghan, Gautham Krishna Gudur, Mohit Malu, Edison Thomaz, Giulia Pedrielli, Pavan Turaga, Hassan Ghasemzadeh
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.10046v2 Announce Type: replace Abstract: Wearable sensors in Internet of Things (IoT) ecosystems increasingly support applications such as remote health monitoring, elderly care, and smart home automation, all of which rely on robust human activity recognition (HAR). Continual learning sy...
169. Transformers Can Learn Rules They've Never Seen: Proof of Computation Beyond Interpolation ​
Author: Andy Gray
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.17019v2 Announce Type: replace Abstract: A central question in the debate over large language models is whether transformers can learn rules they have never seen, or whether they can only interpolate: predict new cases from their similarity to training examples. We test this in a controll...
170. SENSE: Efficient EEG-to-Text via Privacy-Preserving Semantic Retrieval ​
Author: Akshaj Murhekar, Christina Liu, Abhijit Mishra, Shounak Roychowdhury, Jacek Gwizdka
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.17109v2 Announce Type: replace Abstract: Decoding brain activity into natural language is a major challenge in AI with important applications in assistive communication, neurotechnology, and human-computer interaction. Most existing Brain-Computer Interface (BCI) approaches rely on memory...
171. Compressed Video Aggregator: Content-driven Module for Efficient Micro-Video Recommendation ​
Author: Yang Xiao, Huiyuan Chen, Kaiyuan Deng, Chao Jiang, Zinan Ling, Ruimeng Ye, Fei Wang, Xiaolong Ma, Bo Hui
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.08810v2 Announce Type: replace Abstract: We propose \textbf{Compressed Video Aggregator} (CVA), a lightweight micro-video recommendation module that decouples video information from preference learning. CVA first summarizes frozen VFM frame embeddings into a semantic-consensus anchor thro...
172. Exact Symmetry as Algebra: A Machine-Verified Tensor Calculus that Enforces Physical Selection Rules ​
Author: Paulina Hoyos, Shashanka Ubaru, Dongsung Huh, Vasileios Kalantzis, Kenneth L. Clarkson, Misha Kilmer, Haim Avron, Lior Horesh
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.RA
arXiv:2605.20440v2 Announce Type: replace Abstract: Symmetry is central to the physical sciences, yet machine learning usually captures it only approximately, leaving a residual per-step equivariance error $\varepsilon$ that compounds with depth $M$ as $M\varepsilon$, whereas exact equivariance hold...
173. When One Point Is Not Enough: Addressing Ambiguous Instances in Dimensionality Reduction by Splitting ​
Author: Diede P. M. van der Hoorn, Alessio Arleo, Fernando V. Paulovich
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.23540v2 Announce Type: replace Abstract: Dimensionality Reduction (DR) methods are widely used to visualize high-dimensional data. One key task in DR-based analysis is discovering neighborhoods, which relies on analyzing the fine-grained local structure of a projection. However, DR is an ...
174. Towards Verifiable Transformers: Solver-Checkable Circuit Explanations ​
Author: Neel Somani
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.LO
arXiv:2605.24033v2 Announce Type: replace Abstract: Mechanistic interpretability typically discovers circuits and then argues what they do from examples and ablations. We introduce Verifiable Transformers, a framework for turning task-localized circuits into bounded, solver-checkable claims: project...
175. When Rule Violations Are Rare: Chimera Training for Logical Anomaly Detection ​
Author: Alejandro Ascarate, Leo Lebrat, Rodrigo Santa Cruz, Clinton Fookes, Olivier Salvado
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.26171v2 Announce Type: replace Abstract: Many practical anomalies are not merely rare inputs, but violations of semantic constraints: objects co-occur in structured ways, actions imply preconditions, and events satisfy temporal or relational regularities. We study anomaly detection in thi...
176. Mapping small reservoirs across Brazil from 1984 to 2025 ​
Author: Kylen Solvik, Luis Gustavo Carvalho, Marcia N. Macedo
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.00675v2 Announce Type: replace Abstract: Water research in Brazil largely overlooks the widespread damming of small streams for agricultural uses including watering cattle, farm-scale hydropower, irrigation, and aquaculture. These ubiquitous dams and their reservoirs affect water temperat...
177. Target localization, identification and sensing using latent symmetries ​
Author: David Dukov, Malte R"ontgen, Bryn Davies
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.01421v2 Announce Type: replace Abstract: We show that an array of scatterers which has been designed to have latent ("hidden") symmetries can be used as a sensor. We use the capacitance matrix as a canonical model for three-dimensional hybridisation and study how the introduction of an "i...
178. An Empirical Audit of Input Encoders for Multi-Channel Signal Transformers ​
Author: Ossi Lehtinen
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.04752v3 Announce Type: replace Abstract: Transformers consuming multi-channel scalar signals must embed $C$ simultaneous values into one $d_{\text{model}}$-dimensional vector per time step. We audit eight input encoders -- a shared-scalar baseline, per-channel linear projections, an ortho...
179. SafeECGMatch: Calibration-Aware Joint Frequency and Time Space Semi-Supervised Learning for Open-Set ECG Classification ​
Author: Hongkyu Koh, Ikbeom Jang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.08037v2 Announce Type: replace Abstract: Electrocardiogram (ECG) classification models often suffer from severe label scarcity, making semi-supervised learning (SSL) an attractive strategy for reducing annotation costs. In clinical settings, however, unlabeled pools frequently contain out...
180. Projective Graph Residualization: Variation-Allocation Frontiers for Control-Function IV ​
Author: Rui Wu, Zongyuan Chen, Hong Xie, Defu Lian, Enhong Chen
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.14636v2 Announce Type: replace Abstract: Control-function instrumental-variable estimators pass an estimated first-stage residual to an outcome model. The residual must retain the latent control direction while leaving enough treatment variation to identify the structural effect. These de...
181. From Approximation to Emergence: A Theory of Deep Learning ​
Author: Zhilin Zhao
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.01311v2 Announce Type: replace Abstract: Deep learning has outgrown any single mathematical explanation. From Approximation to Emergence develops a unified, proof-oriented account of modern deep learning theory, tracing a path from the classical foundations of approximation, optimization,...
182. GEqTrain: A Configuration-Driven Framework for Retargeting Equivariant Graph Neural Networks Across 3D Scientific Tasks ​
Author: Daniele Angioletti, Marco Nobile, Vittorio Limongelli
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM
arXiv:2607.19083v2 Announce Type: replace Abstract: Equivariant graph neural networks provide a powerful modeling language for three-dimensional scientific data, but their reuse is often limited by implementations tied to specific tasks, outputs, and training regimes. We present GEqTrain, a configur...
183. REGEN: Replay-recycling for Expert-to-Generalist distillation with Offline Reinforcement Learning ​
Author: Yunjie Chen, Xiaoxin Chen, Fang Wang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19450v2 Announce Type: replace Abstract: Large-scale online reinforcement learning (RL) is the predominant means of eliciting advanced abilities including long-term reasoning and agentic tool use in large language models (LLMs). However, continuing to scale it across vast task domains of ...
184. Adaptive Multi-Horizon Reinforcement Learning ​
Author: Manoosh Samiei, Doina Precup, Paul Masset
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.20656v3 Announce Type: replace Abstract: Effective decision-making in complex and changing environments requires balancing short-term and long-term consequences. In reinforcement learning (RL), this trade-off is typically controlled through a fixed discount factor, which imposes a single ...
185. On the Depth Scalability of Logic Gate Networks ​
Author: Taegun An, Dohun kim, Haebeom Lee, Changhee Joo
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.LO
arXiv:2607.21633v2 Announce Type: replace Abstract: Logic Gate Networks (LGNs) compute through compositions of Boolean operations, yet existing LGNs do not reliably benefit from increased depth. We identify two causes: optimization collapse and topology-induced degradation of output-specific credit ...
186. Measuring the Dependency Gap: Diagnosing Inter-Column Fidelity in Tabular Generative Models ​
Author: Jie Zhang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.21636v2 Announce Type: replace Abstract: Synthetic tabular data is prized for preserving not just each column's marginal distribution but the dependencies between columns - structure that carries much of the discriminative signal for minority classes in imbalanced domains such as fraud de...
187. Learning What Matters: Supervising Global Context Pruning with Causal Evidence Sets ​
Author: Jim Allchin
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.21692v2 Announce Type: replace Abstract: Sparse attention prunes a long context to the blocks a model needs, and the usual selector is distilled from a dense teacher's attention. That assumes attention shows which context the answer depends on. We test the assumption on retrieval tasks wh...
188. CausalGate: Causal Importance Distillation for Transformer Module Pruning ​
Author: Kiran Nair, Smriti Regmi, Rodrigue Rizk
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CV, stat.ML
arXiv:2607.22720v2 Announce Type: replace Abstract: Existing adaptive inference methods for Large Language Models rely on observational heuristics, such as hidden-state similarity or activation magnitudes, to drop redundant modules. However, these correlation-based metrics often fail to capture subt...
189. LC-SEPLM: long-range contact-supervised adaptation for sequence-only protein representation learning ​
Author: Chen Wang, Boming Kang, Qinghua Cui
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.22777v2 Announce Type: replace Abstract: Protein language models learn transferable sequence representations. However, because they primarily model contextual dependencies along amino-acid sequences, their training objectives do not explicitly constrain the model to learn three-dimensiona...
190. Directional Influence Function: Estimating Training Data Influence in Constrained Learning ​
Author: Xin Wang (Jeff), R. Tyrrell Rockafellar (Jeff), Xuegang (Jeff), Ban
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.23388v3 Announce Type: replace Abstract: As constrained learning becomes increasingly common, models are trained under explicit feasibility requirements to enforce fairness, safety, robustness, regulariza- tion, and physics or logic constraints. Understanding how training samples in- flue...
191. A Comparison of Data Augmentation Methods for Training Deep Neural Networks on Synthetic Aperture Sonar ​
Author: C. J. Moore, Gregory D. Vetaw, Jordan Malof
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.23770v3 Announce Type: replace Abstract: In this work we study Automatic Target Recognition (ATR) for Synthetic Aperture Sonar (SAS) data with a focus on deep neural networks (DNNs). The main challenge in training DNNs for SAS-ATR arises from the limited quantity of labeled target example...
192. Why does Greedy Search produce Optimal Clustering Outcomes? A Fixed-Core Assignment Theory ​
Author: Kaifeng Zhang, Kai Ming Ting, Sanjay Chawla
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.24237v2 Announce Type: replace Abstract: Many existing clustering methods are designed based on a set-oriented definition---a cluster is a set of similar points---relying a point-to-point similarity function to find similar points. This works well for compact clusters, but clustering perf...
193. Interpretable GOHR Agents via Sparse Autoencoders ​
Author: Shiwei Tan, Yusong Zhao, Weiyi Qin, Wentian Wang, Jacob Feldman, Lazaros K. Gallos, Paul B. Kantor, Vladimir Menkov, Hao Wang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.25132v2 Announce Type: replace Abstract: A central challenge in interpreting learned decision-making systems is to determine whether their internal representations contain concepts that help explain their behavior. We report interpretability experiments for a tokenized autoregressive Tran...
194. Tools Are Not Islands: Set-Level Tool Retrieval for LLM Agents via Query-Conditioned Hyperedge Prediction ​
Author: Xinyi Hong, Pinjun Dong, Xinyang Yu, Binyan Jiang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR
arXiv:2607.25718v2 Announce Type: replace Abstract: Large language model (LLM) agents increasingly rely on invoking external tools to complete real-world tasks. Tool retrieval, which selects a small task-relevant subset from a library of thousands of tools before the agent acts, has therefore become...
195. Differentially Private Permutation Tests ​
Author: Ilmun Kim, Antonin Schrab
Published: 7/30/2026, 4:00:00 AM
Categories: math.ST, cs.CR, cs.LG, stat.ME, stat.ML, stat.TH
arXiv:2310.19043v3 Announce Type: replace-cross Abstract: Recent years have witnessed growing concerns about the privacy of sensitive data. In response to these concerns, differential privacy has emerged as a rigorous framework for privacy protection, gaining widespread recognition in both academic ...
196. One-Frame Calibration with Siamese Network in Facial Action Unit Recognition ​
Author: Shuangquan Feng, Virginia R. de Sa
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2409.00240v2 Announce Type: replace-cross Abstract: Automatic facial action unit (AU) recognition is used widely in facial expression analysis. Most existing AU recognition systems aim for cross-participant non-calibrated generalization (NCG) to unseen faces without further calibration. Howeve...
197. Physics-Informed Graph Neural Networks for Robust AC-Optimal Power Flow ​
Author: Anna Varbella, Damien Briens, Blazhe Gjorgiev, Giuseppe Alessio D'Inverno, Priya L. Donti, Giovanni Sansavini
Published: 7/30/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2410.04818v2 Announce Type: replace-cross Abstract: We present PINCO, an unsupervised learning framework that integrates Graph Neural Networks with physics-informed neural networks for AC optimal power flow (AC-OPF) solutions. Unlike state-of-the-art unsupervised methods that require prescreen...
198. Learning Controlled Stochastic Differential Equations ​
Author: Luc Brogat-Motte, Riccardo Bonalli, Alessandro Rudi
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2411.01982v2 Announce Type: replace-cross Abstract: We study the problem of learning controlled stochastic differential equations (SDEs) [ dX_t = b(t,X_t,u_t),dt + \sigma(t,X_t,u_t),dW_t, ] whose drift and diffusion depend nonlinearly on time, state, and control values. From trajectory dat...
199. Financial Volatility and Risk Forecasting Incorporating a Larger Number of Realized Measures ​
Author: Qianli Zhao, Chao Wang, Richard Gerlach, Giuseppe Storti, Lingxiang Zhang
Published: 7/30/2026, 4:00:00 AM
Categories: q-fin.RM, cs.LG, econ.EM
arXiv:2411.17136v2 Announce Type: replace-cross Abstract: Realised volatility has become increasingly prominent in volatility forecasting due to its ability to capture intraday price fluctuations. With a growing variety of realised volatility estimators, each with unique advantages and limitations, ...
200. Optimal Causal Annotations: An Application to Casenotes in Social Services ​
Author: Ezinne Nwankwo, Lauri Goldkind, Angela Zhou
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.CY, cs.LG, econ.EM, stat.ME
arXiv:2502.10605v4 Announce Type: replace-cross Abstract: Problem definition: Estimating causal effects of interventions is central to policy and operations, but outcome data are often missing or costly to obtain. LLMs can provide text annotation at scale but may be subject to unknown bias. When gro...
201. Low-cost Embedded Breathing Rate Determination Using 802.15.4z IR-UWB Hardware for Remote Healthcare ​
Author: Anton Lambrecht, Stijn Luchie, Jaron Fontaine, Ben Van Herbruggen, Adnan Shahid, Eli De Poorter
Published: 7/30/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2504.03772v3 Announce Type: replace-cross Abstract: Respiratory diseases account for a significant portion of global mortality. Affordable and early detection is an effective way of addressing these ailments. To this end, a low-cost commercial off-the-shelf (COTS), IEEE 802.15.4z standard comp...
202. MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent ​
Author: Hongli Yu, Tinghong Chen, Jiangtao Feng, Jiangjie Chen, Weinan Dai, Qiying Yu, Ya-Qin Zhang, Wei-Ying Ma, Jingjing Liu, Mingxuan Wang, Hao Zhou
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2507.02259v2 Announce Type: replace-cross Abstract: Despite improvements by length extrapolation, efficient attention and memory modules, handling infinitely long documents with linear complexity without performance degradation during extrapolation remains the ultimate challenge in long-text p...
203. HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring ​
Author: Xin Wang, Ting Dang, Xinyu Zhang, Vassilis Kostakos, Michael J. Witbrock, Hong Jia
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.HC, cs.LG
arXiv:2509.07260v5 Announce Type: replace-cross Abstract: Mobile and wearable healthcare monitoring play a vital role in facilitating timely interventions, managing chronic health conditions, and ultimately improving individuals' quality of life. Previous studies on large language models (LLMs) have...
204. Breaking the Stealth-Potency Trade-off in Clean-Image Backdoors with Generative Trigger Optimization ​
Author: Binyan Xu, Fan Yang, Di Tang, Xilin Dai, Kehuan Zhang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.CR, cs.LG
arXiv:2511.07210v3 Announce Type: replace-cross Abstract: Clean-image backdoor attacks, which use only label manipulation in training datasets to compromise deep neural networks, pose a significant threat to security-critical applications. A critical flaw in existing methods is that the poison rate ...
205. DP-MGTD: Privacy-Preserving Machine-Generated Text Detection via Adaptive Differentially Private Entity Sanitization ​
Author: Lionel Z. Wang, Yusheng Zhao, Jiabin Luo, Xinfeng Li, Lixu Wang, Yinan Peng, Haoyang Li, XiaoFeng Wang, Wei Dong
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG
arXiv:2601.04641v2 Announce Type: replace-cross Abstract: The deployment of Machine-Generated Text (MGT) detection systems necessitates processing sensitive user data, creating a fundamental conflict between authorship verification and privacy preservation. Standard anonymization techniques often di...
206. $\texttt{AMEND++}$: Benchmarking Eligibility Criteria Amendments in Clinical Trials ​
Author: Trisha Das, Mandis Beigi, Jacob Aptekar, Jimeng Sun
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2601.06300v2 Announce Type: replace-cross Abstract: Clinical trial amendments frequently introduce delays, increased costs, and administrative burden, with eligibility criteria being the most commonly amended component. We introduce \textit{eligibility criteria amendment prediction}, a novel N...
207. Physics-Informed Singular-Value Learning for Cross-Covariances Forecasting in Financial Markets ​
Author: Efstratios Manolakis, Christian Bongiorno, Rosario Nunzio Mantegna
Published: 7/30/2026, 4:00:00 AM
Categories: q-fin.ST, cs.LG, stat.ML
arXiv:2601.07687v3 Announce Type: replace-cross Abstract: Recent advances in nonlinear shrinkage yield asymptotically optimal cleaners for large covariance matrices and have been extended to empirical cross-covariances via singular-value shrinkage. However, these approaches rely on stationarity and ...
208. The Rise of AI in Weather and Climate Information and its Impact on Global Inequality ​
Author: Amirpasha Mozaffari, Amanda Duarte, Lina Teckentrup, Stefano Materia, Gina E. C. Charnley, Lluis Palma, Eulalia Baulenas Serra, Dragana Bojovic, Paula Checchia, Aude Carreric, Francisco Doblas-Reyes
Published: 7/30/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.AI, cs.LG
arXiv:2603.05710v2 Announce Type: replace-cross Abstract: AI development's current trajectory risks automating and amplifying the North-South divide in the global climate information system. Frontier models are built almost exclusively in the Global North, and this inequality continues through input...
209. Persistence Spheres: a Bi-continuous Linear Representation of Measures for Partial Optimal Transport ​
Author: Matteo Pegoraro
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.AT, math.ST, stat.TH
arXiv:2603.15384v2 Announce Type: replace-cross Abstract: We improve and extend persistence spheres, introduced in~\cite{pegoraro2025persistence}. Persistence spheres map an integrable measure $\mu$ on the upper half-plane, including persistence diagrams (PDs) as counting measures, to a function $S(...
210. TabPFN Extensions for Interpretable Geotechnical Modelling ​
Author: Taiga Saito, Yu Otake, Daijiro Mizutani, Stephen Wu
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CE, cs.LG
arXiv:2603.21033v3 Announce Type: replace-cross Abstract: Geotechnical site characterisation relies on sparse, heterogeneous borehole data, where uncertainty quantification and interpretability matter as much as predictive accuracy. We evaluate TabPFN~\citep{Hollmann2025}, a tabular foundation model...
211. Ontology-driven personalized information retrieval for XML documents ​
Author: Ounnaci Iddir, Ahmed-ouamer Rachid, Tai Dinh
Published: 7/30/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2603.21139v2 Announce Type: replace-cross Abstract: This paper addresses the challenge of improving information retrieval from semi-structured eXtensible Markup Language (XML) documents. Traditional information retrieval systems (IRS) often overlook user-specific needs and return identical res...
212. Adaptively Robust LLM Monitoring via Activation Watermarking ​
Author: Toluwani Aremu, Daniil Ognev, Samuele Poppi, Nils Lukas
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CY, cs.LG
arXiv:2603.23171v3 Announce Type: replace-cross Abstract: Providers monitor deployed large language models (LLMs) to detect misuse that they cannot prevent. LLM monitoring is deterministic and often openly available, so $\emph{adaptive}$ attackers with a local copy can search offline for prompts tha...
213. REAP: Automatic Curation of Coding Agent Benchmarks from Interactive Production Usage ​
Author: Smriti Jha, Matteo Paltenghi, Chandra Maddila, Vijayaraghavan Murali, Shubham Ugare, Satish Chandra
Published: 7/30/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2604.01527v4 Announce Type: replace-cross Abstract: Production deployment of AI coding agents requires fast, reproducible evaluation signals. Existing industrial practices trade off speed and fidelity: online A/B testing takes weeks and risks user experience, shadow deployment yields signals t...
214. Graph Signal Diffusion Models for Wireless Resource Allocation ​
Author: Yigit Berkay Uslu, Samar Hadou, Shirin Saeedi Bidokhti, Alejandro Ribeiro
Published: 7/30/2026, 4:00:00 AM
Categories: eess.SP, cs.IT, cs.LG, math.IT
arXiv:2604.05175v2 Announce Type: replace-cross Abstract: We consider constrained ergodic resource optimization in wireless networks with graph-structured interference. We train a diffusion model policy to match expert conditional distributions over resource allocations. By leveraging a primal-dual ...
215. Shot-based quantum encoding: a data-loading paradigm for quantum neural networks ​
Author: Basil Kyriacou, Viktoria Patapovich, Maniraman Periyasamy, Alexey Melnikov
Published: 7/30/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG
arXiv:2604.06135v2 Announce Type: replace-cross Abstract: Efficient data loading remains a bottleneck for near-term quantum machine learning. Existing schemes (angle, amplitude, and basis encoding) either underuse the exponential Hilbert-space capacity or require circuit depths that exceed the coher...
216. Global monitoring of methane point sources using deep learning on hyperspectral radiance measurements from EMIT ​
Author: Vishal V. Batchu, Michelangelo Conserva, Alex Wilson, Anna M. Michalak, Varun Gulshan, Philip G. Brodrick, Andrew K. Thorpe, Christopher V. Arsdale
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, physics.ao-ph
arXiv:2604.10094v2 Announce Type: replace-cross Abstract: Anthropogenic methane (CH4) point sources are critical drivers of near-term climate forcing, safety hazards, and system-inefficiencies. Space-based imaging spectroscopy is an emerging tool for identifying emissions globally, but existing appr...
217. MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval ​
Author: Shaden Alshammari, Kevin Wen, Abrar Zainal, Mark Hamilton, Navid Safaei, Sultan Albarakati, William T. Freeman, Antonio Torralba
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.DL, cs.IR, cs.LG
arXiv:2604.18584v2 Announce Type: replace-cross Abstract: Mathematical problem solving remains a challenging test of reasoning for large language and multimodal models, yet existing benchmarks are limited in size, language coverage, and task diversity. We introduce MathNet, a high-quality, large-sca...
218. Optimality of Sub-network Laplace Approximations: New Results and Methods ​
Author: Swarnali Raha, Kshitij Khare, Rohit K Patra
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2605.09075v2 Announce Type: replace-cross Abstract: Although the Laplace approximation offers a simple route to uncertainty quantification in deep neural networks, its reliance on inverting large Hessian matrices has motivated a range of computationally feasible low-dimensional or sparse appro...
219. A nonlinear extension of parametric model embedding for dimensionality reduction in parametric shape design ​
Author: Andrea Serani, Giorgio Palma, Matteo Diez
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CE, cs.LG, cs.NA, math.NA
arXiv:2605.11759v2 Announce Type: replace-cross Abstract: Dimensionality reduction is essential in simulation-based shape design, where high-dimensional parameterizations hinder optimization, surrogate modeling, and systematic design-space exploration. Parametric Model Embedding (PME) addresses this...
220. HYVINT: Intensity-Driven Hypergraph Generation with Variational Embeddings ​
Author: Xinyi Hong, Shuntuo Xu, Zhou Yu
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2605.16836v2 Announce Type: replace-cross Abstract: Hypergraphs provide a principled framework for modeling polyadic interactions, with applications in recommendation systems, social networks, and molecular modeling. Hypergraph generation remains challenging because incidence structures are di...
221. The Ghost Couple: Correlated LLM Name Priors and Their Haunting of the Web and Academic Publishing ​
Author: Micha{\l} Brzozowski, Neo Christopher Chung
Published: 7/30/2026, 4:00:00 AM
Categories: cs.DL, cs.LG
arXiv:2606.02184v2 Announce Type: replace-cross Abstract: These names do not exist. Elena Vasquez and Marcus Chen have appeared as volcano experts, astronauts, thriller protagonists, podcast hosts, and academic co-authors across hundreds of independently produced AI-generated documents, never having...
222. The Score Hamiltonian: Mapping Diffusion Models to Adiabatic Transport ​
Author: Peter Halmos, Boris Hanin
Published: 7/30/2026, 4:00:00 AM
Categories: math-ph, cs.AI, cs.LG, math.MP, physics.data-an
arXiv:2606.05217v4 Announce Type: replace-cross Abstract: We exhibit an exact correspondence between sampling with score-based diffusion models and adiabatic transport of ground states for a family of Schr"odinger operators we call Score Hamiltonians, built from the learned score's quantum potentia...
223. TLA-Prover: Verifiable TLA+ Specification Synthesis via Preference-Optimized Low-Rank Adaptation ​
Author: Eric Spencer, Arslan Bisharat, Brian Ortiz, Khushboo Bhadauria, Mujtaba Nazari, TaiNing Wang, George K. Thiruvathukal, Konstantin Laufer, Mohammed Abuhamad
Published: 7/30/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG, cs.LO
arXiv:2606.06133v4 Announce Type: replace-cross Abstract: TLA+ is a formal specification language for verifying distributed systems and safety-critical protocols. Large language models (LLMs) frequently produce TLA+ specifications that fail the TLC model checker for semantic reasons. Across 25 LLMs,...
224. Minimax-Optimal Generalization Bounds for Smooth Deep Neural Networks Trained by (Stochastic) Gradient Descent ​
Author: Junyu Zhou, Puyu Wang, Dennis Wagner, Yunwen Lei, Marius Kloft, Yiming Ying
Published: 7/30/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2606.06772v2 Announce Type: replace-cross Abstract: Characterizing the optimization dynamics and statistical performance of over-parameterized deep neural networks (DNNs) remains a central challenge in understanding the remarkable success of deep learning. We establish quantitative bounds show...
225. BrainG3N: A Dual-Purpose Tokenizer for Controllable 3D Brain MRI Generation ​
Author: Max Van Puyvelde, Ibrahim Gulluk, Wim Van Criekinge, Olivier Gevaert
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG
arXiv:2606.19651v2 Announce Type: replace-cross Abstract: Three-dimensional (3D) brain MRI is central to clinical neurology and neuro-oncology, where generative models could augment under-represented cohorts, simulate disease trajectories, and support privacy-preserving data sharing. Latent diffusio...
226. MLVC: Multi-platform Learned Video Codec for Real-World Deployment ​
Author: Tanel P"arnamaa, Martin Lumiste, Ardi Loot, Evgenii Indenbom, Andrei Znobishchev, Ando Saabas
Published: 7/30/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG
arXiv:2606.28027v2 Announce Type: replace-cross Abstract: Neural video codecs have surpassed classical codecs in coding efficiency but remain impractical for deployment due to cross-platform incompatibility and high computational cost. Existing quantization-based solutions fail to produce determinis...
227. Revealing Hidden Model Behaviors with Task-Specific Self-Reports ​
Author: Taras Kutsyk, Bartosz Zieli'nski
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.03640v2 Announce Type: replace-cross Abstract: Fine-tuning can give a language model a hidden behavior--it may give false answers under a narrow condition, or give harmful advice only when a prompt touches a particular topic. We introduce the Stabilized Adapter for self-Report (SAR), a li...
228. Teaching Tiny VLA Models Where to Look and How to Move ​
Author: Iok Tong Lei, Ying Jie Yap, Wei Huang, Qingchen Xie, Qianzhi Li, Yujie Zhang, Xiaolong Liu, Zhidong Deng
Published: 7/30/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.04171v3 Announce Type: replace-cross Abstract: Tiny Vision-Language-Action models are appealing for real-time robotic control, but reducing model scale often weakens two capabilities essential for manipulation: task-conditioned spatial grounding and coherent action generation. We introduc...
229. Enhanced Seam Segmentation for Automated Welding Robot in Construction Through Transfer Learning: Addressing Limitations of Bilateral Segmentation Network ​
Author: Keonvin Park, Yong Ann Voeurn, Hyeokjun Kweon, Doyun Lee
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.06150v3 Announce Type: replace-cross Abstract: Reliable seam segmentation is essential for autonomous robotic welding in construction, where harsh illumination, specular reflections, and thin weld geometries often degrade segmentation performance. This study proposes a reflection-robust s...
230. When Top-K Misses the Decision: Tool-Call Drift in Multi-Teacher On-Policy Distillation ​
Author: Jiabin Shen, Guang Chen, Chengjun Mao
Published: 7/30/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.07050v3 Announce Type: replace-cross Abstract: Top-K teacher logits make on-policy distillation tractable, but probability mass is not the same as decision support. In a two-teacher tool-use setting, vanilla generalized knowledge distillation raises tool-call recall while also calling on ...
231. Analyzing Image Encoder Choices and Graph Homophily in GCN Frameworks for Breast Ultrasound Classification ​
Author: Sabahattin Mert Daloglu, Ceren Coskun, Harvey Castro, Soner Hacihaliloglu, Ilker Hacihaliloglu
Published: 7/30/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2607.12054v3 Announce Type: replace-cross Abstract: Breast ultrasound is widely used for screening, yet automated analysis remains challenging due to speckle noise, acquisition variability, and weak separation of benign and malignant cases in standard ultrasound imaging. Graph convolutional ne...
232. ResearchArena: Evaluating Sabotage and Monitoring in Automated AI R&D ​
Author: Lena Libon, Ben Rank, Jehyeok Yeon, David Schmotz, Jeremy Qin, Daniel Donnelly, Derck Prinzhorn, Maksym Andriushchenko
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG
arXiv:2607.19321v2 Announce Type: replace-cross Abstract: As AI agents begin to automate AI R&D, we need ways to assess whether their outputs are safe to deploy, even when the agents themselves may be untrusted. AI control offers one such approach: rather than trusting the agent, it treats it as a p...
233. Relaxed activation analysis of dataflow networks - A clock calculus for machine learning and real-time scheduling ​
Author: William Gaudelier, Albert Cohen, Dumitru Potop Butucaru
Published: 7/30/2026, 4:00:00 AM
Categories: cs.PL, cs.LG
arXiv:2607.21797v2 Announce Type: replace-cross Abstract: Previous work has shown that the simple dataflow primitives of the Lustre language allow the natural, semantically unambiguous, and compact representation of machine learning (ML) applications, including models featuring complex conditional e...
234. An Unofficial FastLAS Tutorial: A Programmer's Guide ​
Author: Fabio Aurelio D'Asaro
Published: 7/30/2026, 4:00:00 AM
Categories: cs.LO, cs.AI, cs.LG
arXiv:2607.23557v2 Announce Type: replace-cross Abstract: FastLAS is a scalable system for Inductive Logic Programming (ILP): you give it some background knowledge, a language bias, and a set of examples, and it searches for a set of logic program rules (a hypothesis) that explains the examples. The...
235. Agent-UCT: Upper Confidence Bounds Applied to Trees for Agentic Workflow Optimization with Cost-Awareness ​
Author: Yang Li, Hai Liu, Dian Shao, Yu Wang, Xiyu Chen, Sergey Volkov, Bozhi Wang, Ziyu Sun, Sihang Liu, Ye Luo, Xiaowei Zhang
Published: 7/30/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.24162v2 Announce Type: replace-cross Abstract: Optimizing agentic workflows, such as retrieval-augmented generation (RAG) pipelines, requires navigating a combinatorial space of discrete component choices under tight evaluation budgets. Existing approaches - heuristic search, black-box op...
236. LLM-based Source Code Compression via Thresholded Symbol Ranking ​
Author: Angelo Nardone, Paolo Ferragina
Published: 7/30/2026, 4:00:00 AM
Categories: cs.IT, cs.CL, cs.LG, math.IT
arXiv:2607.24192v2 Announce Type: replace-cross Abstract: We study the problem of lossless compression of source code, motivated by the storage demands of large-scale software archives, such as Software Heritage (https://www.softwareheritage.org/). General-purpose compressors (e.g., zstd, bzip2) off...
237. Beyond "What to Retrieve": Uncertainty in Retrieval-Augmented Code Generation ​
Author: Chandan Kumar Sah, Li Zhang, Xiaoli Lian
Published: 7/30/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CL, cs.LG
arXiv:2607.24884v2 Announce Type: replace-cross Abstract: Repository-level code generation relies on heterogeneous evidence whose relevance, compatibility, and completeness are inherently uncertain. Similar-code examples, repository context, and project-specific APIs may provide complementary inform...