arXiv cs.LG - 2026-07-16 ​
237 items collected.
1. Automatic Differentiation from Scratch: How PyTorch Computes Gradients in Physics-Informed Neural Networks ​
Author: Abdeladhim Tahimi
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.MS, cs.NA, math.NA
arXiv:2607.13042v1 Announce Type: new Abstract: This paper traces, with explicit numerical values, how PyTorch's automatic differentiation (AD) engine computes gradients for Physics-Informed Neural Network (PINN) training -- a setting that requires two levels of differentiation: computing the physic...
2. Beyond Backbone Backpropagation: A Decoupled Strategy for Efficient Transfer Learning ​
Author: Daniel Vila-Cruz, Laura Mor'an-Fern'andez, Ver'onica Bol'on-Canedo
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2607.13043v1 Announce Type: new Abstract: Deep learning models achieve state-of-the-art image classification but face deployment challenges due to computational costs and energy demands. We propose a lightweight training strategy that adapts normalization layers of the model to the new domain ...
3. Federated Explainable Artificial Intelligence: Roles, Architectures, Evaluation, and Open Challenges ​
Author: Masoume Gholizade, Fabrizio Ruffini, Pietro Ducange, Francesco Marcelloni
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13045v1 Announce Type: new Abstract: Federated Learning (FL) has emerged as a key paradigm for privacy-preserving collaborative model training across distributed and heterogeneous data sources. By keeping raw data local, FL addresses data confidentiality concerns, yet it does not resolve ...
4. What Your Model Threw Away and Why You'll Want It Back: Masking, Fingerprinting, and Privacy from Discarded Geometry ​
Author: Zachary P. Bradshaw
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, math.RT, stat.ML
arXiv:2607.13046v1 Announce Type: new Abstract: We develop a framework for the information discarded by machine learning models whose inputs carry a Lie group action. Given a representation $\pi$ of a Lie group $G$ on a space $V$ and a learned function $f\colon V \to \mathbb{R}$, we define two objec...
5. Targeted Recovery of Weight-Space Mechanisms From Neural Networks ​
Author: Antoine Vigouroux, Lee Sharkey
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13047v1 Announce Type: new Abstract: Parameter decomposition (PD) decomposes neural networks into interpretable computational components that faithfully reflect the original network's operations. However, scaling PD to large models requires vast compute, making it a costly and risky endea...
6. Uncertainty-Aware Sequential Decision Rules for Event-Triggered LLM Invocation in Streaming Systems ​
Author: Zhaohui Wang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.13048v1 Announce Type: new Abstract: Streaming inference pipelines increasingly pair lightweight fast models with Large Language Models (LLMs) that provide rich semantic understanding at substantial cost. The central question of when to invoke the LLM has received limited formal treatment...
7. TSSM: Triaxial State Space Model for Global Station Weather Forecasting with Temporal-Variable-Historical Modeling ​
Author: Songru Yang, Zili Liu, Tao Han, Ben Fei, Fenghua Ling, Lei Bai, Chang Liu, Xiangyang Ji, Zhenwei Shi, Zhengxia Zou
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13101v1 Announce Type: new Abstract: Global Station Weather Forecasting (GSWF) is pivotal for localized and extreme weather prediction over key regions. Despite efforts to exploit look-back windows, existing methods show limited accuracy gains and struggle with extreme events and error ac...
8. Disentangling Knowledge States with Ability and Proficiency Modeling for Knowledge Tracing ​
Author: Duantengchuan Li, Yingqian Bi, Jinsong Chen, Rui Zhang, Mingwen Tong
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13103v1 Announce Type: new Abstract: Knowledge tracing (KT) aims to predict students' future performance by modeling their evolving knowledge states from historical interactions. Existing KT methods usually treat the raw interaction sequence as a unified behavioral process, overlooking th...
9. STKAN: Kolmogorov-Arnold Networks for Spatio-Temporal Forecasting ​
Author: Sicong Lai, Yuehong Hu, Siru Zhong, Si Qiao, Yuxuan Liang, Guangyin Jin
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13108v1 Announce Type: new Abstract: Real-world traffic data exhibit heterogeneous spatial correlations and nonlinear temporal dynamics, posing substantial challenges for accurate spatio-temporal forecasting. Existing approaches have developed increasingly sophisticated graph, attention, ...
10. A Hybrid Mamba for Audio-Visual Navigation ​
Author: Yi Wang, Yinfeng Yu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP
arXiv:2607.13110v1 Announce Type: new Abstract: Since the paradigm centered on convolutional neural networks and recurrent architectures was established in 2020, the fundamental backbone networks for audio-visual navigation have undergone no essential changes for more than five years, making them in...
11. CoDiffGRN: Rethinking Gene Regulatory Network Inference via the BEELINE-KGC Benchmark and Co-evolutionary Discrete Diffusion ​
Author: Jiaze Song, Runhao Zhao, Minghao Xu, Bin Cui, Wentao Zhang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13120v1 Announce Type: new Abstract: Inferring gene regulatory networks (GRNs) from single-cell transcriptomic data is crucial for biological discovery, yet existing approaches suffer from a fundamental misalignment with real-world needs. Researchers typically seek a small set of high-con...
12. ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation ​
Author: Qingyu Zhang, Qianhao Yuan, Hongyu Lin, Yaojie Lu, Xianpei Han, Le Sun, Xiang Li, Ming Xu, Jiarui Li, Xiuyin Zhao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.13124v1 Announce Type: new Abstract: Structured pruning is a hardware-friendly way to compress LLMs, but it is mostly validated on multiple-choice recognition tasks, while the same compressed checkpoints can collapse on the free-form generation that deployment actually requires. Two obser...
13. HEDGEHOG: Hierarchical Evaluation of Drug Generators Through Rigorous Filtration ​
Author: Daria A. Ryabchenko (Ligand Pro, Moscow, Russia, Skolkovo Institute of Science and Technology, Artificial Intelligence Center, Moscow, Russia), Pavel Gurevich (Ligand Pro, Moscow, Russia, Skolkovo Institute of Science and Technology, Artificial Intelligence Center, Moscow, Russia), Shamil Kadyrov (Ligand Pro, Moscow, Russia), Daria Frolova (Ligand Pro, Moscow, Russia, Skolkovo Institute of Science and Technology, Artificial Intelligence Center, Moscow, Russia), Kseniia Fedisheva (Ligand Pro, Moscow, Russia), Sergei A. Nikolenko (Ligand Pro, Moscow, Russia), Alexander Shapeev (Ligand Pro, Moscow, Russia, Skolkovo Institute of Science and Technology, Artificial Intelligence Center, Moscow, Russia), Marina A. Pak (Ligand Pro, Moscow, Russia)
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.SE
arXiv:2607.13155v1 Announce Type: new Abstract: Generative molecular models can support early drug discovery by proposing new candidate compounds de novo. In practice, useful candidates must balance target-relevant activity, synthetic accessibility, physicochemical properties, and other multiparamet...
14. SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy ​
Author: Yassine Chemingui, Chenhua Fan, Honghao Wei, Janardhan Rao Doppa
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13175v1 Announce Type: new Abstract: Safe reinforcement learning typically enforces safety by bounding expected cumulative costs, a criterion that often fails to detect rare but catastrophic tail events. To overcome these limitations, this paper introduces SteinGate, a boundary-aware dist...
15. Concurrent Image Understanding and Generation: Self-Correcting Coupled Markov Jump Processes ​
Author: Minh-Quan Le, Armand Comas, Alexandros Lattas, Stylianos Moschoglou, Pedro V'elez, Amit Raj, Aaron Germuth, Thabo Beeler, Dimitris Samaras, Di Qiu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13188v1 Announce Type: new Abstract: Human cognition does not separate understanding and generation. A teacher at a whiteboard speaks and draws $\textit{together}$, each modality reshapes the other. In this paper, we bring this coupled loop to artificial systems. Masked Diffusion Models (...
16. EMAGN: Efficient Multi-Attention Graph Network via Learned Clustering for Scalable Traffic Forecasting ​
Author: Mingxing Xu, Rakesh Chowdary Machineni, Ke Liu, Xi Cheng, Chengqi Lu, Xin Hu, Lyuhao Chen, Xiangyu Li, Junwei You, Oliver Gao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13241v1 Announce Type: new Abstract: Traffic forecasting is highly challenging due to complex and nonlinear spatial and temporal dependencies. Self-attention mechanisms have been widely adopted to model dynamic and long-range dependencies, achieving state-of-the-art performance, but suffe...
17. Reassessing Muon for Matrix Factorization ​
Author: Ali Parviz, Gal Mishne, Alex Cloninger
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13246v1 Announce Type: new Abstract: Muon has recently emerged as a strong optimizer for large-scale deep learning, where it reshapes gradient updates through approximate orthogonalization and has been reported to outperform Adam and AdamW in large language model training. Its empirical s...
18. Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners ​
Author: Haseeb Shah, Lingwei Zhu, Adam White, Martha White
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13274v1 Announce Type: new Abstract: Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discovery and drinking water treatment, where reliability is essential and tuning budgets are limited. Actor-...
19. Tabular Foundation Models for Discrete Choice Estimation ​
Author: Liu Liu, Dan Zhang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, econ.EM
arXiv:2607.13314v1 Announce Type: new Abstract: Tabular foundation models (TFMs) generate predictions on structured data via in-context learning, without task-specific estimation. We ask whether TFMs can be effectively applied to discrete choice, a central demand estimation framework in marketing an...
20. Accuracy-Preserving Stability Regularization for Large-Scale Retail Demand Forecasting ​
Author: Jize Li, Jiani He, Dishu Yang, Dingyan Shang, Jingjing Liu, Shiqi Huang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13331v1 Announce Type: new Abstract: Retail demand forecasts are reused across replenishment, capacity, labor, and transportation planning cycles. Point-error objectives do not constrain abrupt movement between adjacent forecasts, while post-hoc smoothing acts only after model fitting. We...
21. Agora: Collective and Permissionless Internet-Scale Pretraining of Large Language Models ​
Author: Gil Avraham, Violetta Shevchenko, Hadi Mohaghegh Dolatabadi, Karol Pajak, James Snewin, Harry Xi, Rodney O'Donnell, Thalaiyasingam Ajanthan, Sameera Ramasinghe, Chamin Hewa Koneputugodage, Shamane Siriwardhana, Alexander Long
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2607.13332v1 Announce Type: new Abstract: Training large language models at the multi-billion to trillion parameter scale is confined to datacenters, where data-parallel (DP) and model-parallel (MP) techniques presume homogeneous accelerators, high-speed interconnects, and a single orchestrati...
22. Weight Feedback Computes the Jacobian Transpose Locally in Modern Deep Networks ​
Author: Junlong Shen, Xingyu Li
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13380v1 Announce Type: new Abstract: Predictive Coding (PC) offers a biologically motivated alternative to backpropagation via local weight updates, yet routing error between layers still relies on an autograd Jacobian-transpose ($J^\top$) product - the last non-local operation in PC. We ...
23. Where Should RL Post-Training Compute Go? Model Size, Search, Learning, and Feedback ​
Author: Patrick Wilhelm, Odej Kao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.13389v1 Announce Type: new Abstract: Reinforcement Learning (RL) post-training is increasingly used to adapt foundation models for reasoning, planning, and feedback-driven robot-learning pipelines, but constrained post-training resources are often summarized by a single total FLOP budget....
24. Self-Improving is Often Sudden: Enlightenment-style Finetuning for Large-Scale Models ​
Author: Jing-Xiao Liao, Tianwei Zhang, Yu-Hao Jiang, Feifei Zhang, Hang-Cheng Dong, Feng-Lei Fan
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13395v1 Announce Type: new Abstract: The pursuit of autonomously self-improving models has attracted growing interest in the era of large-scale foundation models. Drawing inspiration from the concept of "enlightenment" or "aha moment" in human brain, we hypothesize that large models exhib...
25. Is the Statistical Advantage Worth the Cost? An Empirical Comparison of KANs and MLPs for Structured Data Classification ​
Author: Matthew Steven P. Toledo, Justine Raphael H. Jacinto, Vivekjeet Singh Chambal, Rodolfo C. Camaclang III, Jamlech Iram N. Gojo Cruz, Reginald Neil C. Recario
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13413v1 Announce Type: new Abstract: This study presents an empirical benchmarking comparison between Kolmogorov-Arnold Networks (KANs) and Multi-Layer Perceptrons (MLPs) on structured tabular classification tasks. Motivated by the growing interest in KANs as an alternative function-appro...
26. EXPLORE: Exploration with Guided Search for Analog Topology Generation using Language Models ​
Author: Guanglei Zhou, Chen-Chia Chang, Yikang Shen, Jonathan Ku, Isaac Jacobson, Jingyu Pan, Yiran Chen, Xin Zhang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13416v1 Announce Type: new Abstract: Automating analog circuit topology design is essential to reduce the extensive manual effort required to meet increasingly diverse and customized application demands. Recent advances have applied sequence-to-sequence fine-tuning on pretrained language ...
27. OrDA: Orthogonal Disentanglement of Access Habits Framework for Homepage Marketing Block Recommendations ​
Author: Lingxiao Zhang, Xiaobo Li, Tao Xu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13420v1 Announce Type: new Abstract: Clicks on homepage marketing blocks are driven by a dual-mechanism of content interest and access habits. However, habitual clicks often create Pseudo-Positives in marketing slots, where position advantage masks mediocre content quality, leading to bia...
28. Temperature Scaling Is Not Enough: Calibration Gaps Under Human Label Distributions ​
Author: Wisdom Dogah
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13423v1 Announce Type: new Abstract: Temperature scaling is the dominant post-hoc calibration method in modern deep learning. Its theoretical justification rests on an assumption that is rarely stated explicitly: that ground-truth labels are one-hot and deterministic. In practice, labels ...
29. Data-Efficient Adaptation of LLMs via Attention Head Reweighting ​
Author: Tuomas Oikarinen, Zixiao Chen, Charlotte Siska, Tsui-Wei Weng, Chandan Singh, Jianfeng Gao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.13425v1 Announce Type: new Abstract: Learning effectively from limited data is critical in domains like security where labeled examples are scarce. Large language models (LLMs) have demonstrated some capabilities for data-efficient learning, especially through parameter-efficient adaptati...
30. PUe: Biased Positive-Unlabeled Learning Enhancement by Causal Inference ​
Author: Xutao Wang, Hanting Chen, Tianyu Guo, Yunhe Wang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13428v1 Announce Type: new Abstract: Positive-Unlabeled (PU) learning aims to achieve high-accuracy binary classification with limited labeled positive examples and numerous unlabeled ones. Existing cost-sensitive-based methods often rely on strong assumptions that examples with an observ...
31. Discrete Diffusion Models: A Unified Framework from Tokenization to Generation ​
Author: Ye Yuan, Weien Li, Rui Song, Zeyu Li, Haochen Liu, Xiangyu Kong, Zixuan Dong, Linfeng Du, Zipeng Sun, Weixu Zhang, Jiaxin Huang, Changjiang Han, Yonghan Yang, Zichen Zhao, Xiuyuan Hu, Haolun Wu, Yankai Chen, Fengran Mo, Jikun Kang, Bowei He, Philip S. Yu, Xue Liu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.13431v1 Announce Type: new Abstract: Discrete denoising diffusion models (DDMs) have recently emerged as a compelling alternative to autoregressive (AR) modeling for discrete data, offering parallel generation and iterative global refinement capabilities. Unlike continuous diffusion, wher...
32. Local Redundancy: An Information-Theoretic Measure of Plasticity from Synthetic Memorization ​
Author: Jiaxuan Cheng
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13432v1 Announce Type: new Abstract: Plasticity -- a neural network's ability to adapt to new tasks -- is critical for continual and transfer learning. Existing measures, such as effective rank, dead neuron fraction, and weight norm, lack theoretical grounding and correlate poorly with pe...
33. Distributionally Robust and Safe Imitation Learning ​
Author: Ahmed Aboudonia, Naira Hovakimyan
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2607.13436v1 Announce Type: new Abstract: Imitation learning (IL) has achieved remarkable success in complex decision-making tasks. However, its performance is highly sensitive to distribution shifts, which can pose significant safety risks. We propose a distributionally robust and safe IL fra...
34. PQFA: Parallel Quantum Feature Augmentation of Fused Representations for Multimodal Classification ​
Author: Mingzhu Wang, Yun Shang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, quant-ph
arXiv:2607.13466v1 Announce Type: new Abstract: Most multimodal learning methods improve how heterogeneous representations are aligned and fused, while post-fusion enhancement remains less explored. We propose Parallel Quantum Feature Augmentation (PQFA), a hybrid quantum-classical framework that ap...
35. Explainable Artificial Intelligence for Anomaly Detection in Banking Transactions: An Internal Audit Perspective ​
Author: Anupa Lodhi
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13469v1 Announce Type: new Abstract: The banking sector increasingly relies on automated systems to monitor electronic transactions for signs of fraud, yet conventional rule-based approaches struggle with high false-positive rates and offer no justification for their outputs, limiting the...
36. DeepLoop: Depth Scaling for Looped Transformers ​
Author: Shuzhen Li, Yifan Zhang, Jiacheng Guo, Quanquan Gu, Mengdi Wang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13491v1 Announce Type: new Abstract: Looped Transformers scale sequential computation by applying a compact stack of physical blocks for multiple rounds, increasing unrolled depth without increasing stored parameters. This reuse changes the residual-scaling problem: in an untied Transform...
37. A VAE-Driven Multi-Task Satellite-Aided Semantic Communication Framework for 6G-Enabled Connected Autonomous Vehicles ​
Author: S. M. Abtahiul Alam, Niloy Das, Apurba Adhikary, Yu Qiao, Zhu Han, Choong Seon Hong
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2607.13494v1 Announce Type: new Abstract: The development of smart transportation systems and the introduction of 6G wireless communication technologies have significantly changed vehicle network topologies. Future connected autonomous vehicle (CAV) networks require bandwidth-efficient, reliab...
38. Factorized Spectral Representations for Reinforcement Learning ​
Author: Junyi Wu, Dan Li
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13498v1 Announce Type: new Abstract: Learning a compact model of the world from interaction data is central to sample-efficient deep reinforcement learning. Spectral representation methods have become the leading paradigm for representation learning in continuous control by taking a matri...
39. CDS: Counterfactual Directionality Score for Structured Interventions in Spatial Graphs ​
Author: Humaira Anzum, Md Ishtyaq Mahmud, Jagan Mohan Reddy Dwarampudi, Tania Banerjee
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13508v1 Announce Type: new Abstract: Quantifying directional influence between node populations is a fundamental problem in graph-based modeling, particularly in spatial biological systems where cell-cell interactions shape functional outcomes. Existing approaches based on attention, attr...
40. ExTernD: Expanded-Rank Ternary Decomposition Ternary LLM PTQ with Accuracy Approaching Any Quantization Level ​
Author: Chethan Reddy G. P
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13511v1 Announce Type: new Abstract: We introduce ExTernD (Expanded-rank Ternary Decomposition), a post-training factorization of each LLM weight matrix $A \in \mathbb{R}^{m \times n}$ into $A \approx B \mathrm{diag}(D) C$ with ternary factors $B \in {-1,0,+1}^{m \times k}$, $C \in {-1...
41. Clustering algorithms for multivariate wind farm SCADA data filtering ​
Author: Nicol`o Italiano, Vasilis Pettas, Tuhfe G"o\c{c}men, Nicolaos A. Cutululis
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13544v1 Announce Type: new Abstract: During wind farm operation, Supervisory Control and Data Acquisition (SCADA) systems record numerous anomalies, transients, and specific operational modes, leading to large datasets. However, for a wide range of applications, only measurements correspo...
42. From Novice to Expert: Cost-Aware Bandits for Evolving Worker Performance in Crowdsensing ​
Author: Yin Huang, Qingsong Liu, Jie Xu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13546v1 Announce Type: new Abstract: Mobile crowdsensing (MC) recruits mobile users to perform sensing tasks using their smartphones, enabling large-scale applications such as traffic monitoring and environmental sensing. A fundamental challenge is online worker recruitment under uncertai...
43. Structured Reinforcement Learning for Bayesian Persuasion : Application to Intelligent Interactive Driving ​
Author: Merlin Paul, Anup Aprem
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2607.13576v1 Announce Type: new Abstract: Interactive driving, wherein an intelligent lead vehicle equipped with real-time traffic data coordinates route choices of connected vehicles, offers a promising approach to dynamic traffic management. To address the challenge of harmonising decisions,...
44. Gauge-Invariant, Parameter-Insensitive Regularization for Potential Recovery from Flow on Directed Graphs ​
Author: Mohammad Forouhesh
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.IR, eess.SP, stat.ML
arXiv:2607.13609v1 Announce Type: new Abstract: Recovering a latent potential from observed flow on a directed graph (a discrete Poisson problem with Dirichlet boundaries) is ill-posed, and the standard fix backfires: ridge regularization shrinks toward a gauge-meaningless origin, collapsing and rev...
45. The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models ​
Author: Fabio Arnez, Alexandra Gomez-Villa
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13612v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) are the dominant design for latent world models, yet they are usually justified by empirical performance rather than a normative principle. We show that the choice of anti-collapse regulariser determines...
46. FastCentNN: Accelerating Centroid Neural Network with Entropy Proxy ​
Author: Le-Anh Tran
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.13613v1 Announce Type: new Abstract: Centroid neural network (CentNN) is an unsupervised competitive learning algorithm in which centroid splitting is triggered only after strict local stabilization, often leading to prolonged low-movement training phases before model expansion. This repo...
47. How the Hessian-Spectrum of Neural Networks Depends on Data ​
Author: Jasraj Singh, Enea Monzio Compagnoni, Antonio Orvieto
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13631v1 Announce Type: new Abstract: The Hessian matrix is an important quantity of interest when it comes to studying the loss landscape and optimization dynamics in deep learning, as well as designing measures of generalization, second-order learning algorithms, etc. Prior works have fo...
48. Consensus as Privileged Context for Label-Free Self-Distillation ​
Author: John Gkountouras, Josip Juki'c, Ivan Titov
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.13643v1 Announce Type: new Abstract: Sampling multiple solutions and returning the majority answer is among the most reliable ways to improve the reasoning accuracy of large language models without labels, and a growing family of methods converts this consensus signal into training superv...
49. Maximally Robust Satisficing Bayesian Optimization ​
Author: Samuli Kinnunen, Petrus Mikkola, Antti Niskanen, Arto Klami
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13652v1 Announce Type: new Abstract: Many design tasks can be cast as black-box function optimization, enabling use of Bayesian optimization to find an ideal design with minimal number of trials. However, often we do not actually need the optimum but instead a sufficiently good solution i...
50. The Hyperspherical Geometry of CLIP Latent Space: A Semantic Mixture Model ​
Author: Zijie Yu, Gaowen Liu, Ramana Rao Kompella, Philip S. Yu, Yue Song
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13660v1 Announce Type: new Abstract: Contrastive Language-Image Pretraining (CLIP) representations form a semantic embedding space governed by cosine similarity, reflecting an intrinsic hyperspherical geometry. However, existing probabilistic interpretations typically rely on Gaussian ass...
51. Optimal and Efficient Contextual Combinatorial Semi-bandits with General Function Approximation ​
Author: Hao Qin, Chicheng Zhang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13686v1 Announce Type: new Abstract: We study the contextual combinatorial semi-bandit (CCSB) problem with general reward function approximation. At each round, the learner observes a context, selects a combinatorial action consisting of a subset of basic arms, and receives the reward of ...
52. Microstructure-Conditioned Surrogate Models for Graded Multiscale Optimization of Mycelium Composites ​
Author: J. Storm, I. B. C. M. Rocha, S. Schyck, K. Masania, F. P. van der Meer
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.NA, math.NA, physics.comp-ph
arXiv:2607.13688v1 Announce Type: new Abstract: Emerging sustainable materials increasingly rely on engineered hierarchy and microstructure to achieve control of their properties and mechanical behavior. Optimizing these materials with controllable microstructures requires efficient multiscale simul...
53. Conditional Invertible Neural Networks for Data-Driven UAV Control: A 2-D Proof of Concept ​
Author: Christian Wittke, Stephan Myschik, Oliver Niggemann
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2607.13703v1 Announce Type: new Abstract: We investigate conditional invertible neural networks (cINNs) as probabilistic inverse-dynamics models for multirotor control. For a planar X8 coaxial multicopter, we learn $p(u \mid s_t, c_t)$ from an incremental nonlinear dynamic inversion (INDI) tea...
54. DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention ​
Author: Xing Lei, Wenyan Yang, Xuetao Zhang, Donglin Wang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.13731v1 Announce Type: new Abstract: Goal-conditioned reinforcement learning hinges on how the goal is encoded. Contrastive, metric, temporal-distance, and information-theoretic encoders differ in objective. They still share one trait. None of them sees the current state. Such a state-ind...
55. Constraint-Driven Model Optimization: An Industry Framework for Selecting Compression and Acceleration Techniques in Modern Machine Learning Systems ​
Author: Dhruv Shivkant, Saket Mohanty, Utkarsh Wadhwa
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13735v1 Announce Type: new Abstract: The rapid deployment of machine learning systems across cloud, edge, and enterprise environments has brought model optimization to the forefront of systems-engineering. Despite a rich literature spanning quantization, pruning, knowledge distillation, p...
56. Implementations of Quantum and Classical Topology-Aligned Architectures for Molecular Property Prediction ​
Author: James T. Pegg, Hubert Okadome Valencia, Ronin Wu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13737v1 Announce Type: new Abstract: For low-data and resource-constrained regimes typical of quantum chemistry, parameter-efficient learning is a key objective. Here, we propose a topology-aligned inductive bias in which the model architecture mirrors the molecular bond graph: atoms map ...
57. Algebraic Representability as the Limiting Regime of Grokking: An Exactly Solvable Model with Holomorphic Activations ​
Author: Chon-Fai Kam, Xavier Cadet, Miloud Bessafi, Frederic Cadet
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.13749v1 Announce Type: new Abstract: Neural networks trained on modular arithmetic exhibit grokking, a delayed transition from memorisation to generalisation known to depend on model capacity: too little and the network memorises slowly or not at all, too much and it generalises almost im...
58. MxGPS: Multiplex Graph Transformers for a Power Grid Foundation Model ​
Author: Charilaos Papaioannou, Ioannis Tsantilas, Dimitris Giannakakos, Vasilis Michalakopoulos, Sotiris Pelekis, Vangelis Marinakis, Arsam Aryandoust, Antonello Monti, Ricardo J. Bessa, Perdo P. Vergara, Jochen Cremer, Elissaios Sarmas
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13763v1 Announce Type: new Abstract: Single-task fine-tuning of graph neural networks (GNNs) for power grid problems exhibits a systematic failure mode: models that achieve the lowest in-distribution error degrade the most under topology shift. We term this topology overfitting: the tende...
59. Mono-Z Dark Matter Search with Neural Spline Flows Using CMS Run 2015D Open Data ​
Author: Hitesh Rasineni (VIT-AP University, Amaravati, India), Bhavishya Chebrolu (Mohan Babu University, Tirupati, India)
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, hep-ex
arXiv:2607.13771v1 Announce Type: new Abstract: We report a search for dark matter (DM) produced in association with a leptonically decaying (Z) boson at (\sqrt{s}=13) TeV using CMS Run 2015D open data corresponding to an integrated luminosity of (2.32,\mathrm{fb}^{-1}) together with simplifi...
60. NodeImport: Imbalanced Node Classification with Node Importance Assessment ​
Author: Nan Chen, Zemin Liu, Bryan Hooi, Bingsheng He, Jun Hu, Jia Chen
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.13837v1 Announce Type: new Abstract: In real-world applications, node classification on graphs often faces the challenge of class imbalance, where majority classes dominate training, resulting in biased model performance. Traditional GNNs often struggle in such scenarios, as they tend to ...
61. Heavy-Tailed Flow Matching via Random Clocks ​
Author: Zhouhao Yang, Yezhen Wang, Kenji Kawaguchi, Vladimir Braverman, Haoyang Cao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.13841v1 Announce Type: new Abstract: Heavy-tailed data arise in many domains where rare events carry disproportionate importance, such as imbalanced image datasets, financial returns, and weather extremes. Standard diffusion and flow-matching models typically begin from Gaussian noise or ...
62. Relevance-Aware Rule: Structural Deletion of Irrelevant Conditions in Decision Trees ​
Author: Jung-Sik Hong, Jeongeon Lee, Min Kyu Sim, Sangheum Hwang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13874v1 Announce Type: new Abstract: Decision trees generate interpretable if--then rules, yet they contain irrelevant conditions (IRCs). These IRCs arise from the structural mechanism of tree splitting and persist even in modern optimal sparse tree induction algorithms. Existing IRC dele...
63. AI-Augmented Adaptive Digital Twin Modeling for Brain Tumor Evolution Prediction and Treatment Scheduling ​
Author: Wenxi Liu, Michael Trimboli, Xianqi Li
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13877v1 Announce Type: new Abstract: Brain tumor progression exhibits spatially heterogeneous growth, patient-specific treatment response, and complex interactions with surrounding anatomy, making accurate long-term prediction challenging. We propose an AI-augmented adaptive digital twin ...
64. Task-Oriented Sensing and Covert Transmissions for Collaborative Multi-AUV Systems ​
Author: Xueyao Zhang, Chenyang Yan, Bo Yang, Xuelin Cao, Zhiwen Yu, Bin Guo, George C. Alexandropoulos, Merouane Debbah, Chau Yuen
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13880v1 Announce Type: new Abstract: In underwater covert cooperative missions, autonomous underwater vehicles (AUVs) often cannot rely on active sonar to continuously obtain complete information, since active sensing and frequent communications increase the risk of exposure. As a result,...
65. PiVoT: A Variational Solution for Real-time Large-scale Multi-object Detection and Tracking under Heavy Clutter ​
Author: Runze Gan, Qing Li, Simon J. Godsill, Mike E. Davies, James R. Hopgood
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, eess.SP, stat.ML
arXiv:2607.13891v1 Announce Type: new Abstract: Multi-object detection and tracking from noisy point clouds remain challenging in many data-scarce radar applications. Current Bayesian trackers based on Poisson measurement models offer a training-free solution but struggle to achieve accuracy and eff...
66. RF Spectrogram Anomaly Detection with Quantum Kitchen Sinks: Architecture, Representation, and Hardware Validation ​
Author: Abdallah Aaraba, Alexis Vieloszynski, Remon Polus, Ola Ahmad, Soumaya Cherkaoui
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13897v1 Announce Type: new Abstract: The broadcast nature of wireless channels exposes radio-frequency (RF) networks to anomalous and malicious transmissions, making anomaly detection a fundamental requirement for secure spectrum management. Quantum Kitchen Sinks (QKS) offer a lightweight...
67. An Efficient Newton Algorithm for Nonnegative Matrix Factorization with the Kullback-Leibler Divergence ​
Author: Damien Lesens, J'er'emy E. Cohen, Bora U\c{c}ar
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13919v1 Announce Type: new Abstract: Nonnegative Matrix Factorization (NMF) is a fundamental tool in unsupervised learning, which approximates a nonnegative matrix by the product of two low-rank nonnegative factors. The Kullback-Leibler (KL) divergence is best suited to measure the data t...
68. VAIOM: Continuous-Input, Discrete-Output Decoder-Only Financial Sequence Modeling ​
Author: Yiming Ma, Xinyu Chen
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP
arXiv:2607.13929v1 Announce Type: new Abstract: Financial observations are continuous, heterogeneous, and noisy, whereas decoder-only next-token models are usually built around discrete symbolic inputs. We introduce Vector-Input Autoregressive Inference for Ordinal-Return Modeling (VAIOM), a decoder...
69. TRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon Agents ​
Author: Leitian Tao, Baolin Peng, Wenlin Yao, Tao Ge, Hao Cheng, Mike Hang Wang, Jianfeng Gao, Sharon Li
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13988v1 Announce Type: new Abstract: Multi-turn agents solve complex tasks through extended sequences of tool interactions before producing a final answer, making credit assignment a fundamental challenge during post-training. Outcome rewards provide reliable supervision for short-horizon...
70. Lyapunov Exponent as Physics-Informed Dense Reward: RL Discovery of Stabilization Beyond the Kapitza Pendulum ​
Author: Slava Andrejev
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14001v1 Announce Type: new Abstract: We suggest using the Lyapunov characteristic exponent (LCE) as a dense reward signal for the reinforcement learning problem of stabilizing the inverted pendulum with vertical motion. With LCE, the agent not only successfully found the oscillatory motio...
71. Lighthouse RL: Sample-Efficient Circuit Optimization via Strategic Reset Points ​
Author: Mustafa Emre G"ursoy, Stefan Uhlich, Ryoga Matsuo, Ya\u{g}{\i}z Gen\c{c}er, Arun Venkitaraman, Chia-Yu Hsieh, Andrea Bonetti, Eisaku Ohbuchi, Lorenzo Servadei
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AR
arXiv:2607.14008v1 Announce Type: new Abstract: In this paper, we introduce Lighthouse RL, a sample-efficient reinforcement learning (RL) approach for analog circuit sizing. Traditional methods lack generalization across different performance targets, while standard RL approaches waste resources exp...
72. Transforming Rank: How Architecture Navigates the Spectral Pathologies of Depth ​
Author: Katie Everett
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14018v1 Announce Type: new Abstract: We investigate how each component of the Transformer feedforward block architecture design determines how much rank survives across depth at initialization. We reinterpret skip connections and normalization, long understood as controlling magnitude, as...
73. Improving Wind and Solar Power Prediction with Efficient Wrapper-based Feature Selection: An Empirical Study ​
Author: Daniel Grillmeyer, Marius Hadry, Michael Stenger, Vanessa Borst, Veronika Lesch, Samuel Kounev
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14024v1 Announce Type: new Abstract: With rising global energy demand and growing awareness of climate change and its impacts, the share of renewable energies in the global energy mix continues to grow. Unlike conventional power generation, the output of renewable energy sources cannot be...
74. MetaPerch: Learning from metadata for bioacoustics foundation models ​
Author: Mustafa Chasmai, Vincent Dumoulin, Jenny Hamer
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.SD
arXiv:2607.14072v1 Announce Type: new Abstract: Bioacoustic foundation models rely on large-scale citizen science platforms like Xeno-Canto for geographically and ecologically diverse data. Recent work has shown that supervision alone can produce SotA species detection models when trained on this la...
75. Linear Independent Component Analysis via Optimal Transport ​
Author: Ashutosh Jha, Michel Besserve, Simon Buchholz
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.14081v1 Announce Type: new Abstract: Linear Independent Component Analysis (ICA) recovers jointly independent source signals from their linear mixtures. To achieve this, classical ICA algorithms attempt to maximize non-Gaussianity, measured by negentropy, which is linked to independence b...
76. Leveraging unlabelled data for generalizable neural population decoding ​
Author: Ximeng Mao, Nanda H. Krishna, Avery Hee-Woon Ryoo, Matthew G. Perich, Guillaume Lajoie
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC
arXiv:2607.14086v1 Announce Type: new Abstract: Robust and accurate neural decoders are integral to neurotechnologies such as brain-computer interfaces and closed-loop experiments. Recent work has shown that tokenizing neural data at the spike level facilitates multi-session pretraining and delivers...
77. Quantum Circuit Vision: Cost-Aware Evaluation of Visual AI Agents for Quantum Code Generation ​
Author: Dongping Liu, Aoyu Zhang, Luyao Zhang
Published: 7/16/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.CV, cs.LG
arXiv:2607.10057v1 Announce Type: cross Abstract: Can AI agents visually comprehend quantum circuit diagrams and generate verified executable code--and at what cost? We present Quantum Circuit Vision, a cost-aware evaluation framework for multimodal AI agents on quantum circuit visual understanding....
78. FixItFlow: Automated Troubleshooting Guide Generation from Cloud Incidents ​
Author: Srihari Unnikrishnan, Jaskaran Singh Walia, Drishti Goel, Supriyo Ghosh
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.SE
arXiv:2607.13035v1 Announce Type: cross Abstract: Cloud services experience frequent incidents that require rapid diagnosis and resolution. Troubleshooting guides help engineers respond consistently, but creating them manually is labor-intensive, resulting in incomplete coverage and outdated documen...
79. LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents ​
Author: Ravidu Suien Rammuni Silva, Ahmad Lotfi, Isibor Kennedy Ihianle, Golnaz Shahtahmassebi, Jordan J. Bird
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG, cs.MM
arXiv:2607.13041v1 Announce Type: cross Abstract: Large Language Model (LLM) based AI educational content generation systems are increasingly being developed, yet no standardised benchmark exists to systematically evaluate them. This study introduces LessonBench-V1, a benchmark dataset comprising 64...
80. Precomputing the Future-Offset Average in TriAttention ​
Author: Amarnath Mukherjee (Hozhoke, Inc.)
Published: 7/16/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2607.13051v1 Announce Type: cross Abstract: TriAttention is a recent method for shrinking the KV cache of long-reasoning LLMs: it scores each cached key by how much attention it is likely to receive and evicts the lowest-scoring ones. Because a key does not know how far away its future queries...
81. HRIBench: Benchmarking Interaction-Centric Human-Robot Collaboration ​
Author: Chang Liu, Jiawei Zhang, Tao Zhang, Ye Wang, Hongyu Zhou, Qin Jin
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.13056v1 Announce Type: cross Abstract: Current vision-language-action (VLA) benchmarks primarily evaluate isolated manipulation skills while leaving human-robot interaction structure largely unmodeled. However, real-world collaboration fundamentally requires coordination under shared agen...
82. When is the combined load identifiable from a stress-intensity profile? A coupled forward-inverse study on SIFBench finite-element data ​
Author: Giansalvo Cirrincione, Filippo Grassia
Published: 7/16/2026, 4:00:00 AM
Categories: math.NA, cond-mat.mtrl-sci, cs.AI, cs.LG, cs.NA
arXiv:2607.13074v1 Announce Type: cross Abstract: This work studies the inverse problem of recovering the relative magnitudes of the tension, bending, and bearing loads acting on a crack from its stress-intensity-factor profile along the crack front, using the public SIFBench finite-element data. Th...
83. The Entanglement Wall: Activation-Space Probes as Risk Detectors, Not Context Adjudicators ​
Author: Dominik Schwarz
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.13075v1 Announce Type: cross Abstract: Context can change whether a request is harmful without changing its topic or surface form. We ask whether residual-stream probes distinguish harmful requests from surface-matched benign controls at a useful operating point. Across three 7-8B model f...
84. Operational Evidence Gaps for LLMs in Fraud Detection and Trust-and-Safety Workflows ​
Author: Keyur Gabani
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.13078v1 Announce Type: cross Abstract: LLMs are now proposed for fraud detection, scam investigation, content moderation, and other trust-and-safety workflows. Much of the public literature still evaluates them as models, with less attention to their behavior as components in operational ...
85. SingGuard-NSFA: Extensible Guardrails for Agentic AI via Generative Reasoning and Real-Time Classification ​
Author: SingGuard Team
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG
arXiv:2607.13081v1 Announce Type: cross Abstract: We present nsfaguard, a guardrail framework for securing agentic AI systems against operational threats, such as prompt injection, sensitive information extraction, malicious code requests, dangerous tool misuse, and resource exhaustion. We first int...
86. Securing LLMs in the Wild: Privacy and Security Challenges at the Edge ​
Author: Ren-Yi Huang, Mingchen Li, Dumindu Samaraweera, Morris Chang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.13088v1 Announce Type: cross Abstract: Large Language Models (LLMs) are rapidly moving from research settings into the wild, deployed on enterprise infrastructure, personal devices, and edge platforms. While cloud deployments offer scalable compute, concerns over data sovereignty, complia...
87. Self-Improvements in Modern Agentic Systems: A Survey ​
Author: Zhe Ren, Yimeng Chen, Dandan Guo, Guowei Rong, Tonghui Li, R. B. Xiong, Qingfeng Lan, Wenyi Wang, Li Nanbo, Yibo Yang, Mingchen Zhuge, J"urgen Schmidhuber
Published: 7/16/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.13104v1 Announce Type: cross Abstract: Self-improving autonomous agents are moving from research prototypes to deployed systems. The primary goal is controllable evolution, or adaptation, from experience with minimal or even no human input. This survey frames modern self-improving agents ...
88. DeepCormack: Fermi surface tomography using model-based data-driven algorithms ​
Author: Georg F. B. Lovric, Bryn Drury, Carola-Bibiane Sch"onlieb, Stephen B. Dugdale, Ander Biguri
Published: 7/16/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cond-mat.str-el, cs.LG
arXiv:2607.13107v1 Announce Type: cross Abstract: The experimental reconstruction of the 3D two-photon momentum density (TPMD) via angular correlation of electron-positron annihilation radiation (ACAR) is a particularly useful method for studying material Fermi surfaces. It does not rely on low temp...
89. Improving Molecular Property Prediction in Small Language Models Using Graph-based Tools ​
Author: Konstantinos Bougiatiotis, Dimitrios Kelesis, Georgios Paliouras
Published: 7/16/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.13115v1 Announce Type: cross Abstract: Small language models (SLMs) have shown promise for zero-shot molecular property prediction from SMILES strings, yet they often suffer from structural blindness because sequence representations under-specify key graph-topological cues. We propose a m...
90. Mixed-Timescale Differential Coding for Downlink Model Broadcast in Wireless Federated Learning ​
Author: Chung-Hsuan Hu, Zheng Chen, Erik G. Larsson
Published: 7/16/2026, 4:00:00 AM
Categories: cs.IT, cs.DC, cs.LG, eess.SP, math.IT
arXiv:2607.13119v1 Announce Type: cross Abstract: In standard federated learning systems, the parameter server broadcasts the global model to the participating devices in every iteration. Motivated by the temporal correlation between consecutive global models, differential coding can be applied to g...
91. Text2Sign: A Single-GPU Diffusion Baseline for Text-to-Sign Language Video Generation ​
Author: Ruize Xia
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.CV, cs.LG
arXiv:2607.13164v1 Announce Type: cross Abstract: Sign language is a primary communication channel for millions of Deaf and hard-of-hearing people, yet text-to-signer video generation remains costly because video diffusion models are expensive to train and evaluate. This paper presents Text2Sign, a ...
92. Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models ​
Author: Ilias Kazantzidis, Timothy J. Norman, Yali Du, Christopher T. Freeman
Published: 7/16/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.13172v1 Announce Type: cross Abstract: We address the problem of safely training an agent policy and deploying a good and safe policy, in settings where the environment dynamics are unknown and no suitable reward function is available. In the context of safety-critical environments, we co...
93. BARS: Benign-Anchored Ranking and Selection for False Alarm Reduction in Network Intrusion Detection ​
Author: Abu Fuad Ahmad, Istiaque Ahmed
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.13203v1 Announce Type: cross Abstract: False alarms remain a major barrier to deploying network intrusion detection systems (NIDS). In high-volume environments, even a sub-1% false positive rate can generate tens of thousands of daily alerts. Filter-based feature selection is attractive b...
94. Adaptive Filtering of the KV Cache: Diagnosing and Correcting Structural-Role Bias in LLM Inference ​
Author: Soumil Mandal
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.13205v1 Announce Type: cross Abstract: Attention-based KV cache eviction (H2O and its descendants) compresses the memory-constrained state of a long-context model by ranking tokens on accumulated attention mass, treated here as signal energy, and keeping the heaviest. On schema-dense inpu...
95. Classifying daily activities needs posture, reconstructing them needs motion ​
Author: Arefeh Farahmandi, Gunnar Blohm
Published: 7/16/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.IR, cs.LG
arXiv:2607.13216v1 Announce Type: cross Abstract: Humans recognize movements effortlessly, even from noisy and complex visual input. But what information in the stimulus allows humans to rapidly classify movements? No framework has systematically compared different strategies of movement analysis to...
96. Hierarchical $\mathcal{F}$-Clustering: Approximation and Hardness of Clustering into Trees and Bounded Diameter Graphs ​
Author: Micha{\l} Szyfelbein, Dariusz Dereniowski
Published: 7/16/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2607.13217v1 Announce Type: cross Abstract: Consider the following variation on the Hierarchical Clustering problem: Usually, while building a hierarchical clustering, one recursively partitions the data until each cluster becomes a singleton. We relax the halting condition of the recursive pr...
97. Graph Partitioning with Demands: Generalized Conductance and its Applications ​
Author: Micha{\l} Szyfelbein, Dariusz Dereniowski
Published: 7/16/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2607.13218v1 Announce Type: cross Abstract: In this work, we study various graph partitioning problems under a general demand model. In each such task, we are given a graph $G=(V,E,c,w)$ with a capacity function $c\colon E\to \mathbb{N}$ and a demand function $w\colon V\times V\to \mathbb{N}$....
98. Audited Selective Verification for Risk-Controlled N-1 Thermal Contingency Screening under Deployment Shift ​
Author: Jayakumar Manoharan
Published: 7/16/2026, 4:00:00 AM
Categories: eess.SY, cs.AI, cs.LG, cs.SY
arXiv:2607.13221v1 Announce Type: cross Abstract: Real-time N-1 contingency screening in an energy management system trades assurance against cost: verifying every credible outage with full power flow is too slow, while fast linear-sensitivity screening gives no statistical guarantee and can silentl...
99. Active Learning for Efficient Annotation of Surgical Videos with Weak Supervision ​
Author: Manasa Dendukuri, Matjaz Jogan, Daniel A. Hashimoto, Guiqiu Liao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.13237v1 Announce Type: cross Abstract: Precise spatial-temporal annotation of laparoscopic videos is time-consuming and requires expert knowledge. We propose a human-in-the-loop knowledge acquisition framework that combines active learning with dual-loss optimization to significantly redu...
100. Theory-Level Autoformalization: From Isolated Statements to Unified Formal Knowledge Bases ​
Author: Marcus J. Min, Mike He, Zhaoyu Li, Zixuan Yi, Sharad Malik, Aarti Gupta, Xujie Si, Osbert Bastani
Published: 7/16/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.PL
arXiv:2607.13292v1 Announce Type: cross Abstract: Autoformalization translates informal natural language into formal, machine-verifiable languages. While most work focuses on individual statements, real formalization efforts are inherently theory-level: they require an entire web of axioms, definiti...
101. Adapting Generalist Vehicle Models for High-Speed MPC Across Terrains ​
Author: Rwik Rana, Jesse Quattrociocchi, Christian Ellis, Nathan Tsoi, Garrett Warnell, Joydeep Biswas
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2607.13319v1 Announce Type: cross Abstract: High-speed off-road autonomy requires precise closed-loop control for a target vehicle while remaining robust across changing terrains. Recent forward kinodynamic (FKD) prediction foundation models suggest a promising path, starting from a generalist...
102. Privacy Preserving Recommender Systems Balancing Personalization with Privacy ​
Author: Ranjeet K Jha, Venkata Suresh Gummadilli
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.13328v1 Announce Type: cross Abstract: Personalized recommendation systems are central to modern e-commerce and retail platforms, but they typically rely on centralized storage of detailed user interaction data, creating significant privacy and regulatory challenges. With increasing requi...
103. Delving into the Temporal Challenges of Unified Video Protection Against Image-to-Video and Fine-Tuning-based Customization ​
Author: Yuxin Huang, Ziming Hong, Mingming Gong, Wanyu Wang, Jing Zhang, Tongliang Liu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.CR, cs.LG
arXiv:2607.13336v1 Announce Type: cross Abstract: Recent diffusion-based video generation models have enabled high-quality personalized video customization through both tuning-based pipelines, which fine-tune a video diffusion model, and reference-based pipelines such as image-to-video generation. H...
104. Evaluation Ability Does Not Imply Optimization Utility: LLM-as-a-Judge Signals in Closed-Loop Table Recognition ​
Author: Donghwan Kim
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.13347v1 Announce Type: cross Abstract: LLM-as-a-judge is widely used to provide feedback and selection signals in closedloop regeneration, but this use remains insufficiently validated. We study it in table recognition, where deterministic TEDS evaluation provides a controlled testbed, us...
105. Learned Pairwise Deep Dual-Optimal Inequalities for Stabilizing Column Generation ​
Author: Zhengzhong Ricky You, Bo Tang, Haoran Liu, Baichuan Mo
Published: 7/16/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2607.13373v1 Announce Type: cross Abstract: Column generation (CG) is central to many large-scale optimization algorithms, including branch-price-and-cut methods for vehicle routing problems, but unstable dual solutions can substantially slow its convergence. Existing deep dual-optimal inequal...
106. GFlowRL: Scaling Distribution-Matching RL to Large Language Models ​
Author: Xiaodong Liu, Michael Xu, Jack W. Stokes, Paul Smolensky, Doug Burger, Jianfeng Gao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.13394v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) offer a promising alternative to reward-maximizing reinforcement learning (RL) for large reasoning models, encouraging diverse reasoning paths by matching reward distributions rather than collapsing to dominant mo...
107. Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations ​
Author: Rui Wang, Hongru Wang, Yi Chen, Boyang Xue, Tianqing Fang, Wenhao Yu, Kam-Fai Wong
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.13399v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a key paradigm in LLM post-training, yet its training dynamics remain poorly understood. We present a systematic study examining the role, pathologies, and regulations of OPD. We first clarify the role of OPD a...
108. Price of Fairness in Bandits: A Tight Minimax Characterization ​
Author: Dhruv Sarkar, Soumyadeep Dutta, Sayak Ray Chowdhury
Published: 7/16/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2607.13402v1 Announce Type: cross Abstract: In bandit problems, standard regret-minimizing algorithms treat exploration as an amortized cost, which can expose early participants to unfair ex-ante losses in settings such as clinical trials. Recent work addresses this by evaluating the sequence ...
109. Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models ​
Author: Chun-Yi Kuan, Siwon Kim, Byeonggeun Kim, Suyoun Kim, Bo-Ru Lu, Qinming Tang, Ankur Gandhe, Hung-yi Lee, Chieh-Chi Kao, Chao Wang
Published: 7/16/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.CL, cs.LG, cs.SD
arXiv:2607.13408v1 Announce Type: cross Abstract: Recent text-to-audio models generate high-quality audio, but often fail to follow instructions involving multiple sound events and temporal order. This gap arises because existing evaluation and training signals mainly emphasize global similarity or ...
110. Evaluating Frontier AI Agents as Autonomous Clinical Security Auditors ​
Author: Michael O. Eniolade
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.13411v1 Announce Type: cross Abstract: Clinical AI models can expose patients to harm when adversarial vulnerabilities go undetected, yet formal security auditing requires statistical expertise, specialized tools, and significant time. We present an open evaluation task, built on METR Tas...
111. Non-Expansive Two-Time-Scale Stochastic Approximation: A Fixed-Schedule One-Quarter Barrier and Bias-Corrected Acceleration ​
Author: Dhruv Sarkar, Vaneet Aggarwal
Published: 7/16/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.13414v1 Announce Type: cross Abstract: Non-expansive two-time-scale stochastic approximation is governed by a slow stochastic Krasnoselskii--Mann fixed-point iteration rather than by contraction to a unique equilibrium. We study this regime under a contractive fast map and a non-expansive...
112. DreamSat-Pose: Spacecraft Pose Estimation from Single-View 3D Reconstructions and Learned 2D-3D Feature Matching ​
Author: Josiane Uwumukiza, Jocelyn Zhao, Giovanni Lavezzi, Giacomo Battaglia, Paolo Panicucci, Minduli C. Wijayatunga, Victor Rodriguez-Fernandez, Richard Linares
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.13449v1 Announce Type: cross Abstract: 6-DoF pose estimation is a critical task in autonomous rendezvous and proximity operations. In the case of an unknown target, this task becomes challenging as it shall be paired with the reconstruction of the target shape model. In this article, we p...
113. HIVE-3D: Hierarchical Voxel Enhancement for High-Quality 3D Scene Generation ​
Author: Bin Zang, Wenting Zheng, Xiaoliang Luo, Zhiyuan Fang, Shi Li, Lvchun Wang, Wei Yu, Yi Zhao, Tian Xie, Yuchi Huo, Rengan Xie
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.13468v1 Announce Type: cross Abstract: Recently, a line of works can generate impressive 3D objects from a single image, but they are limited by restricted representation resolution, making them unsuitable for 3D scene generation. In this work, we introduce HIVE-3D, a novel method for hig...
114. Deformable State Estimation for Autonomous Surgical Tissue Retraction Under Partial Observability ​
Author: Everest Yang, Skye Thompson, George D. Konidaris
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.13475v1 Announce Type: cross Abstract: Surgical tissue retraction requires effective manipulation planning under partial and noisy perception. We study state estimation for deformable tissue retraction, where only sparse observations of the tissue surface are available at decision time. W...
115. Topology-Agnostic Mesh Reconstruction of Deformable Objects from Sparse Touch ​
Author: Everest Yang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.13479v1 Announce Type: cross Abstract: Estimating the full shape of a deformable object is especially challenging when vision is unavailable: in the dark, inside an opaque bag, behind the manipulating hand, or under heavy self-occlusion. Touch is the natural sensor in these settings, but ...
116. When T2I Synthetic Data Backfires: Amplified Privacy Risks in Real-Synthetic Mix Training ​
Author: Na Li, Boyu Kuang, Hongsheng Hu, Liquan Chen, Hyoungshick Kim, Yansong Gao, Anmin Fu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.13541v1 Announce Type: cross Abstract: To overcome data scarcity and privacy constraints in data collection, it has become standard practice across academia and industry to augment real training data with text-to-image (T2I)-generated synthetic data, a paradigm we term Real-Synthetic Mix-...
117. Parallel gradient boosting for flexible estimation of conditional distributions ​
Author: R'emy Chapelle (CESP, CB, EVDG), Nicolas Vayatis (CB), Bruno Falissard (CESP), Mohammed Sedki (CESP)
Published: 7/16/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.13550v1 Announce Type: cross Abstract: Boosting is one of the most successful learning techniques for standard classification and regression tasks. Its extension to multi-output prediction problems has found an increasing number of applications in recent years. Among them is the predictio...
118. Flow-aware Optimal Navigation in Unsteady Flows through Reinforcement Learning ​
Author: Andrea Maria Braghin, Nicol`o Botteghi, Matteo Tomasetto, Andrea Manzoni, Gabriele Cazzulani
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SY, eess.SY
arXiv:2607.13553v1 Announce Type: cross Abstract: Autonomous robotic navigation in nonstationary time-varying fluid flows remains a fundamental challenge due to partial observability and the unpredictability of realistic environments. While classical optimal control frameworks employed in robotics r...
119. Spectral-Informed Neural Networks Outperform Spectral Methods in High-dimensional PDEs ​
Author: Tianchi Yu, Ivan Oseledets
Published: 7/16/2026, 4:00:00 AM
Categories: math.NA, cs.AI, cs.CE, cs.LG, cs.NA
arXiv:2607.13566v1 Announce Type: cross Abstract: For low-dimensional problems ($d\leq3$), spectral methods can achieve exceptionally high accuracy. For middle-dimensional problems ($4 \leq d \lesssim 10$), spectral methods remain feasible through specific techniques such as sparse grids or hyperbol...
120. Graded Entity-Familiarity Readouts in Language Models: Polish Adaptation, Cross-Language Robustness, and Refusal Steering ​
Author: Grzegorz Brzezinka
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.13568v1 Announce Type: cross Abstract: Can a language model estimate its familiarity with an entity before generating an answer? We study activations at the final prompt token in twelve instruction-tuned models from the Bielik, PLLuM, Gemma-4, and Qwen3 families, using a new dataset of 1,...
121. Approximation of solutions of parameter-dependent problems by residual neural networks ​
Author: Ana Carpio
Published: 7/16/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.AP
arXiv:2607.13574v1 Announce Type: cross Abstract: We develop a convergent scheme to train neural networks involving analytic activation functions based on gradient flows. Convergence properties are guaranteed by Lojasiewicz theory. The main advantage of this approach is its simplicity of implementat...
122. Agile perceptive multi-skill locomotion for quadrupedal robots in the wild ​
Author: Jun-Gill Kang, Jaehyun Park, Tae-Gyu Song, Joon-Ha Kim, Seungwoo Hong, Hae-Won Park
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2607.13579v1 Announce Type: cross Abstract: Enabling quadrupedal robots to traverse complex terrains-from rugged outdoor environments to urban landscapes-requires seamless integration of multiple motor skills, smooth transitions between gaits, and high-speed perceptive locomotion using only on...
123. Analogical Deep Research: Retrieving and Integrating Historical Analogies for Foresight Analysis ​
Author: Yongqiang Chen, Guangyi Chen, Yuewen Sun, Kun Zhang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, stat.ML
arXiv:2607.13602v1 Announce Type: cross Abstract: Systematic comparisons between current situations and structurally similar past events in the historical, i.e., historical analogies, is among the most powerful tools for foresight analysis. In this work, we present a new task called Analogical Deep ...
124. STOCKTAKE: Measuring the Gap Between Perception and Action in LLM Agents with a Fair Oracle ​
Author: Sagar Deb, Ashwanth Krishnan
Published: 7/16/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.13618v1 Announce Type: cross Abstract: LLM agents are increasingly evaluated on multi-week decision tasks in which the state that drives cost is never directly observed. On such tasks the final cost cannot say why an agent failed: it may have misread the world, or read it correctly and st...
125. Towards quantum machine learning for assessing the resilience of post-quantum cryptography ​
Author: Jaros{\l}aw A. Miszczak
Published: 7/16/2026, 4:00:00 AM
Categories: quant-ph, cs.CR, cs.LG
arXiv:2607.13722v1 Announce Type: cross Abstract: The potential capabilities of quantum computers motivated the development of cryptographic protocols suitable for securing communication against adversaries with access to large fault-tolerant quantum computers. However, even though current quantum c...
126. Traffic-Aware Randomized Smoothing for LLM-Based Network Intrusion Detection ​
Author: Zhenpeng Li
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.13801v1 Announce Type: cross Abstract: Large language model (LLM)-based intrusion detection systems (IDS) are increasingly studied for security monitoring, yet their robustness against feasible traffic manipulation remains largely empirical. We present Traffic-Aware Randomized Smoothing (...
127. Multimodal Assessment of Pancreatic Cancer Resectability Using Deep Learning ​
Author: Vincent Ochs, Christoph Kuemmerli, Florentin Bieder, Julia Wolleb, Joel L. Lavanchy, Julia Ruppel, Jan Liechti, Stephanie Taha-Mehlitz, Christian Andreas Nebiker, Beat Mueller, Giuseppe Kito Fusai, Joerg-Matthias Pollok, Anas Taha, Philippe C. Cattin, Sebastian Staubli
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.IR, cs.LG
arXiv:2607.13826v1 Announce Type: cross Abstract: Accurate determination of pancreatic ductal adenocarcinoma (PDAC) resectability relies on evaluating how the tumor interacts with major peripancreatic vessels on CT imaging, yet expert assessment often shows substantial variability. We introduce a fu...
128. Quantum Topological Data Encoding ​
Author: Adam Weso{\l}owski, Dimitrios Thanos, Daniel Leykam, Lirand"e Pira
Published: 7/16/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2607.13847v1 Announce Type: cross Abstract: Many datasets encountered across a wide range of domains possess rich geometric and topological structure that is difficult to capture using conventional vector-based representations. Quantum machine learning offers the possibility of processing high...
129. Verifying formulas for interventional distributions ​
Author: Francesco Freni, Leonard Henckel, Sebastian Weichwald
Published: 7/16/2026, 4:00:00 AM
Categories: stat.ME, cs.AI, cs.LG, stat.ML
arXiv:2607.13883v1 Announce Type: cross Abstract: We formalize verification in causal graphical models: deciding whether a given observational formula identifies a target interventional distribution. This opens a problem complementary to identification, asking not whether any identifying formula exi...
130. The 2nd International StepUP Competition for Biometric Footstep Recognition: From Steps to Strides ​
Author: Robyn Larracy, Anant Gupta, Gourav Gupta, Ethan Eddy, Maxime Devanne, Cyril Meyer, Jin-Chern Chiou, Yueh-Shan Lee, Zong-Han Lu, Aaron Tabor, Erik Scheme
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.13905v1 Announce Type: cross Abstract: The International StepUP Competition Series was launched to advance research in pressure-based footstep biometrics through a standardized and challenging evaluation framework. Using the large-scale StepUP-P150 dataset (with more than 200,000 high-res...
131. Partially Correlated Verifier Cascades in LLM Harnesses: Concave Log-Odds, Polynomial Reliability, and Blind-Spot Ceilings ​
Author: Jiangang Han
Published: 7/16/2026, 4:00:00 AM
Categories: math.ST, cs.AI, cs.LG, stat.TH
arXiv:2607.13918v1 Announce Type: cross Abstract: Serial verification gates are a core reliability primitive in LLM harnesses: a candidate answer is returned only if $k$ verifier calls all accept it. Under conditionally independent gates, the recent Odds Law (arXiv:2606.15712) shows that posterior l...
132. Generative Compilation: On-the-Fly Compiler Feedback as AI Generates Code ​
Author: Niels M"undler-Sasahara, Hristo Venev, Dawn Song, Martin Vechev, Jingxuan He
Published: 7/16/2026, 4:00:00 AM
Categories: cs.PL, cs.AI, cs.LG
arXiv:2607.13921v1 Announce Type: cross Abstract: Languages with rich static semantics, such as Rust, provide stronger guarantees for AI-generated code, but their strictness makes generation more difficult. Off-the-shelf compilers can provide useful feedback post-generation, but does not guide inter...
133. Plausible Deniability Guarantees for Whistleblowers ​
Author: Leo Richter, Matt J. Kusner
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, stat.ML
arXiv:2607.13928v1 Announce Type: cross Abstract: Whistleblowers are a key safeguard against organizational wrongdoing, but the threat of retaliation deters reporting. Existing whistleblower-protection proposals lack formal privacy guarantees, and existing differential privacy mechanisms do not dire...
134. A novel unsupervised machine learning strategy to handle multimodal cardiac PET/MRI data ​
Author: Brunnhilde Ponsi (Nantes Universit'e, CHU Nantes, Nantes, France, CRCI2NA, INSERM UMR 1307, Nantes, France), Thomas Carlier (Nantes Universit'e, CHU Nantes, Nantes, France, CRCI2NA, INSERM UMR 1307, Nantes, France), Lara Marteau (Nantes Universit'e, CHU Nantes, Nantes, France, Cardiology Department, INSERM UMR 1307, CIC 1413, l'institut du Thorax, Nantes, France), Aur'elien Monnet (Siemens Healthineers France, Courbevoie, France), Thomas Eug`ene (Nantes Universit'e, CHU Nantes, Nantes, France, CRCI2NA, INSERM UMR 1307, Nantes, France), Jean-Michel Serfaty (Nantes Universit'e, CHU Nantes, Nantes, France, Radiology Department, l'institut du Thorax, Nantes, France), Nicolas Piriou (Nantes Universit'e, CHU Nantes, Nantes, France, Cardiology Department, INSERM UMR 1307, CIC 1413, l'institut du Thorax, Nantes, France), Hatem Necib (Nantes Universit'e, CHU Nantes, Nantes, France, CRCI2NA, INSERM UMR 1307, Nantes, France)
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, physics.med-ph
arXiv:2607.13936v1 Announce Type: cross Abstract: Arrhythmogenic left ventricular cardiomyopathy is a genetic myocardial disease difficult to diagnose due to the lack of gold standard criteria. Simultaneous PET/MR imaging, combined with multiparametric quantitative analysis, could facilitate the ide...
135. Beyond the $d^{2.5}$-mixing bound for Dikin walks on polytopes ​
Author: Yunbum Kook
Published: 7/16/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, math.OC
arXiv:2607.13943v1 Announce Type: cross Abstract: Inspired by interior-point methods (IPM) for structured convex optimization, Kannan and Narayanan introduced the Dikin walk for sampling uniformly from polytopes in 2009. As in IPMs, the Dikin walk is affine-invariant, and its convergence is governed...
136. Multimodal Empirical Bayes Variational Autoencoders for Joint Longitudinal and Time-to-Event Modeling ​
Author: Anders Sj"oberg, Nils Olsson, Marcus Baaz, Mats Jirstrand
Published: 7/16/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.13984v1 Announce Type: cross Abstract: Longitudinal tumor measurements, dropout information, and genetic covariates provide complementary information about treatment response, but integrating these data sources within a single population modeling framework remains challenging. We extend t...
137. Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0 ​
Author: Wenxiao Wang, Priyatham Kattakinda, Soheil Feizi
Published: 7/16/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.14004v1 Announce Type: cross Abstract: Most reported gains from agent-optimization methods are one-shot: an agent is optimized against a fixed benchmark and the resulting improvement is reported as if it were a stable property of the method. This does not test the setting that matters for...
138. Early Adoption of Agentic Coding Tools by GitHub Projects ​
Author: Maliha Noushin Raida, Daqing Hou
Published: 7/16/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CY, cs.LG
arXiv:2607.14037v2 Announce Type: cross Abstract: Agentic coding tools are increasingly capable of generating and submitting pull requests (PRs) to software projects, introducing new forms of human-agent collaboration in software development. While prior studies have examined PR-level outcomes of ag...
139. Multi-Expert Routing for Multi-Domain Low-Resource OCR: A Manchu Case Study ​
Author: Zhan Chen, Jiqiao Ma, Chih-wen Kuo
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.14041v1 Announce Type: cross Abstract: Historical Manchu OCR must accommodate various visually distinct writing styles, including regular script, running script, and the semi-cursive chancery hand used in palace memorials, despite limited labeled data. We study a multi-expert system that ...
140. Screening of Biosecurity Features in Metagenomic Data with Evo 2 Probes ​
Author: Jeremy Guntoro, Alexander Dack, Dylan Danno, Michaela Jan\v{c}ovi\v{c}ov'a, Kri\v{z}an Jurinovi'c, Vanessa Smilansky
Published: 7/16/2026, 4:00:00 AM
Categories: q-bio.GN, cs.LG
arXiv:2607.14070v1 Announce Type: cross Abstract: Genomic foundation models such as Evo 2 learn rich sequence representations, but their value for biosecurity screening is largely unexplored. We ask how much biosecurity-relevant signal is linearly accessible in these representations by training mini...
141. EM-GANSim: Real-time and Accurate EM Simulation Using Conditional GANs for 3D Indoor Scenes ​
Author: Ruichen Wang, Dinesh Manocha
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2405.17366v3 Announce Type: replace Abstract: We present a novel machine-learning (ML) approach (EM-GANSim) for real-time electromagnetic (EM) propagation that is used for wireless communication simulation in 3D indoor environments. Our approach uses a modified conditional Generative Adversari...
142. DarwinLM: Evolutionary Structured Pruning of Large Language Models ​
Author: Shengkun Tang, Oliver Sieberling, Eldar Kurtic, Zhiqiang Shen, Dan Alistarh
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2502.07780v4 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved significant success across various NLP tasks. However, their massive computational costs limit their widespread use, particularly in real-time applications. Structured pruning offers an effective solution ...
143. On the Sublinear Regret of Continuous K-Max Bandits ​
Author: Yu Chen, Siwei Wang, Longbo Huang, Wei Chen
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2502.13467v2 Announce Type: replace Abstract: The $K$-Max combinatorial multi-armed bandit problem arises in applications such as recommendation and distributed decision making, where the reward is determined by the maximum outcome among $K$ selected arms. When outcomes are continuous and only...
144. NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache ​
Author: Donghyun Son, Euntae Choi, Sungjoo Yoo
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2505.18231v3 Announce Type: replace Abstract: Large Language Model (LLM) inference is typically memory-intensive, especially when processing large batch sizes and long sequences, due to the large size of key-value (KV) cache. Vector Quantization (VQ) is recently adopted to alleviate this issue...
145. RF-Informed Graph Neural Networks for Accurate and Data-Efficient Circuit Performance Prediction ​
Author: Anahita Asadi, Leonid Popryho, Inna Partin-Vaisband
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2508.16403v3 Announce Type: replace Abstract: Accurately predicting the performance of active radio frequency (RF) circuits is essential for modern wireless systems but remains challenging due to highly nonlinear behavior and the high computational cost of traditional simulation tools. Existin...
146. Multi-Dictionary Learning for Low Rank Sparse Coding ​
Author: Boya Ma, Abram Magner, Maxwell McNeil, Petko Bogdanov
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.10033v2 Announce Type: replace Abstract: Sparse dictionary coding represents signals as linear combinations of a few dictionary atoms. It has been applied to images, time series, graph signals and multi-way spatio-temporal data by jointly employing temporal and spatial dictionaries. Data-...
147. Representation-Based Exploration for Language Models: From Test-Time to Post-Training ​
Author: Jens Tuyls, Dylan J. Foster, Akshay Krishnamurthy, Jordan T. Ash
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.11686v2 Announce Type: replace Abstract: Reinforcement learning (RL) promises to expand the capabilities of language models, but it is unclear if current RL techniques promote the discovery of novel behaviors, or simply sharpen those already present in the base model. In this paper, we in...
148. Variational Mixture of Graph Neural Experts for Alzheimer's Disease Recognition across Frequency Bands in EEG Brain Networks ​
Author: Jun-En Ding, Anna Zilverstand, Shihao Yang, Albert Chih-Chieh Yang, Feng Liu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.11917v3 Announce Type: replace Abstract: Dementia disorders such as Alzheimer's disease (AD) and frontotemporal dementia (FTD) exhibit overlapping electrophysiological signatures in EEG that challenge accurate diagnosis. Existing EEG-based methods are limited by full-band frequency analys...
149. Computation-aware Energy-harvesting Federated Learning with Pipelined Cyclic Scheduling ​
Author: Eunjeong Jeong, Nikolaos Pappas
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2511.11949v2 Announce Type: replace Abstract: Federated learning (FL) is a powerful paradigm for distributed learning, but increasing model complexity leads to significant energy consumption from client-side computations for local training. This challenge is critical in energy-harvesting FL (E...
150. Lag Operator SSMs: A Geometric Framework for Structured State Space Modeling ​
Author: Sutashu Tomonaga, Kenji Doya, Noboru Murata
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.18965v2 Announce Type: replace Abstract: Structured State Space Models (SSMs), which are at the heart of the recently popular Mamba architecture, are powerful tools for sequence modeling. However, their theoretical foundation relies on a complex, multistage process of continuous-time mode...
151. Mechanistic Evidence for Preserved-but-Misaligned Representations in Non-IID FedAvg ​
Author: Muhammad Haseeb, Salaar Masood, Muhammad Abdullah Sohail, Mohammad Fatim Shoaib, Muhammad Tahir
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.23043v3 Announce Type: replace Abstract: Federated Averaging (FedAvg) often degrades under non-IID client data, but it remains unclear whether this degradation reflects the loss of client-learned representations or a failure to use representations that are still present. We study this que...
152. Survival Dynamics of Neural and Programmatic Policies in Evolutionary Reinforcement Learning ​
Author: Anton Roupassov-Ruiz, Yiyang Zuo
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.04365v2 Announce Type: replace Abstract: In evolutionary reinforcement learning tasks (ERL), agent policies are often encoded as small artificial neural networks (NERL). Such representations lack explicit modular structure, limiting behavioral interpretation. We investigate whether progra...
153. Overcoming the Modality Gap in Context-Aided Forecasting ​
Author: Vincent Zhihao Zheng, 'Etienne Marcotte, Arjun Ashok, Andrew Robert Williams, Lijun Sun, Alexandre Drouin, Valentina Zantedeschi
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.12451v4 Announce Type: replace Abstract: Context-aided forecasting (CAF) holds promise for integrating domain knowledge and forward-looking information, enabling AI systems to surpass traditional statistical methods. However, recent empirical studies reveal a puzzling gap: multimodal mode...
154. Disentangled Latent Dynamics Manifold Fusion for Solving Parameterized PDEs ​
Author: Zhangyong Liang, Huanhuan Gao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.12676v5 Announce Type: replace Abstract: Generalizing neural surrogate models across different PDE parameters remains difficult because changes in PDE coefficients often make learning harder and optimization less stable. The problem becomes even more severe when the model must also predic...
155. Not All Retrievals are Useful: Cross-Attention for Input-Aware RAG in Time Series Forecasting ​
Author: Seunghan Lee, Jaehoon Lee, Jun Seo, Sungdong Yoo, Minjae Kim, Tae Yoon Lim, Dongwan Kang, Hwanil Choi, SoonYoung Lee, Wonbin Ahn
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.14709v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) enhances zero-shot time series (TS) forecasting by leveraging external knowledge bases, yet existing approaches overlook input-level relevance when fusing retrieved samples with the query. We argue that not all ...
156. A Generalised Exponentiated Gradient Approach to Enhance Fairness in Binary and Multi-class Classification Tasks ​
Author: Maryam Boubekraoui, Giordano d'Aloisio, Antinisca Di Marco
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2603.21393v2 Announce Type: replace Abstract: The widespread use of AI and ML models in sensitive areas raises significant concerns about fairness. While the research community has introduced various methods for bias mitigation in binary classification tasks, the issue remains under-explored i...
157. Rethinking Multimodal Fusion for Time Series: Text Modalities Need Constrained Fusion ​
Author: Seunghan Lee, Jun Seo, Jaehoon Lee, Sungdong Yoo, Minjae Kim, Tae Yoon Lim, Dongwan Kang, Hwanil Choi, SoonYoung Lee, Wonbin Ahn
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.22372v3 Announce Type: replace Abstract: Recent advances in multimodal learning have motivated the integration of auxiliary modalities such as text or vision into time series (TS) forecasting. However, most existing methods provide limited gains, often improving performance only in specif...
158. Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies ​
Author: Zhanzhi Lou, Hui Chen, Yibo Li, Qian Wang, Bryan Hooi
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2604.00830v3 Announce Type: replace Abstract: Test-Time Learning (TTL) enables language agents to iteratively refine their performance through repeated interactions with the environment at inference time. At the core of TTL is an adaptation policy that updates the actor policy based on experie...
159. NetForge RL: A Multi-Agent Simulation Environment for Cyber Defense with Durative Actions ​
Author: Igor Jankowski
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.MA
arXiv:2604.09523v3 Announce Type: replace Abstract: Training reinforcement-learning agents for cyber defense requires an environment that reflects the operational setting: noisy, partial observations, several defenders coordinating across a network, and an adaptive adversary realized through self-pl...
160. Partially Observed Structural Causal Models ​
Author: Turan Orujlu, Jordan Matelsky, Martin V. Butz, Charley M. Wu, Konrad P. Kording
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ME, stat.ML
arXiv:2605.03268v2 Announce Type: replace Abstract: Here we introduce Partially Observed Structural Causal Models (POSCMs) as an extension of structural causal models (SCMs) to settings where upstream contexts co-determine both the interaction structure and downstream mechanisms on observed variable...
161. Flatness and Gradient Alignment Are Both Necessary: Spectral-Aware Gradient-Aligned Exploration for Multi-Distribution Learning ​
Author: Aristotelis Ballas, Christos Diou
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2605.07914v2 Announce Type: replace Abstract: Sharpness-aware and gradient-alignment methods have been shown to improve generalization, however each family of methods targets a single geometric property of the loss landscape, while ignoring the other. In this paper, we show that this omission ...
162. Stable Attention Response for Reliable Precipitation Nowcasting ​
Author: Penghui Wen, Zexin Hu, Sen Zhang, Patrick Filippi, Xiaogang Zhu, Allen Benter, Thomas Bishop, Zhiyong Wang, Kun Hu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.13181v2 Announce Type: replace Abstract: Precipitation nowcasting remains challenging due to the highly localized, rapidly evolving, and heterogeneous nature of atmospheric dynamics. Although recent methods increasingly adopt attention-based architectures in both unimodal and multimodal s...
163. Test-Time Learning with an Evolving Library ​
Author: Weijia Xu, Alessandro Sordoni, Chandan Singh, Zelalem Gero, Michel Galley, Xingdi Yuan, Jianfeng Gao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.14477v2 Announce Type: replace Abstract: We introduce EvoLib, a test-time learning framework that enables large language models to accumulate, reuse, and evolve knowledge across problem instances without parameter updates or external supervision. Instead of adapting model parameters, our ...
164. Foundation Models for Credit Risk Prediction: A Game Changer? ​
Author: Bart Baesens, Andreas Goethals, Stefan Lessmann, Simon De Vos, Cristi'an Bravo, David Martens, Victor Medina-Olivares, Christophe Mues, Maria Oskarsd'ottir, Seppe vanden Broucke, Tony Van Gestel, Tim Verdonck, Wouter Verbeke
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.18147v2 Announce Type: replace Abstract: Predictive models play a pivotal role in credit risk management, guiding critical decisions through accurate estimation of default probabilities and losses. Extensive research has introduced new modeling techniques, complemented by large-scale benc...
165. AvAtar: Learning to Align via Active Optimal Transport ​
Author: Qi Yu, Ruizhong Qiu, Zhichen Zeng, My T. Thai, Huan Liu, Hanghang Tong
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.24395v2 Announce Type: replace Abstract: Alignment plays a fundamental role in many machine learning problems, such as multi-network analysis, multimodal learning, and point cloud registration. Recent works increasingly leverage optimal transport (OT) for distributional alignment, whose e...
166. Variational Inference for Evidential Deep Learning ​
Author: Jiawei Tang, Xinyan Du, Hui Liu, Junhui Hou, Yuheng Jia
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.26477v2 Announce Type: replace Abstract: While Deep Neural Networks (DNNs) achieve remarkable performance, their tendency to produce overconfident predictions. Evidential Deep Learning (EDL) mitigates this by formulating predictions as a Dirichlet distribution over class probabilities to ...
167. Algorithmic Recourse of In-Context Learning for Tabular Data ​
Author: Wenshuo Dong, Jiaming Zhang, Shaopeng Fu, Hongbin Lin, Di Wang, Lijie Hu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.31272v2 Announce Type: replace Abstract: As predictive models are increasingly deployed in high-stakes settings such as credit approval, there is a growing need for post-hoc methods that provide recourse to affected individuals. Many such models operate on tabular data, where features cor...
168. Intention Driven Identification of In-Possession Match Phases in Association Football through Temporal Graph Learning ​
Author: Yuesen Li, Daniel Link
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.09289v3 Announce Type: replace Abstract: Understanding tactical organisation in association football requires identifying in-possession match phases that are shaped by evolving tactical intentions rather than by spatial patterns alone. This study proposes an intention-driven framework for...
169. Unlocking Latent Dimensions: Exploring Representations of Large-Scale X-ray Scattering Data using Variational Autoencoders ​
Author: Monika Choudhary, Xiaoya Chong, Runbo Jiang, Wiebke Koepp, Petrus H. Zwart, Damon English, Gregory M. Su, Eric Schaible, Chenhui Zhu, Mostafa Nassr, Noah P. Wamble, Kelvin Kam-Yun Li, Jonathan M. Chan, Jose Carlos Diaz, Cameron McKay, Lynn Katz, Benny Freeman, Guillaume Freychet, Yevgen Matviychuk, Eliot Gann, Daniel B. Allan, Benedikt Sochor, Frank Schluenzen, Stephan V. Roth, Ethan J. Crumlin, Dylan McReynolds, Tanny Chavez, Alexander Hexemer
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.14999v2 Announce Type: replace Abstract: Scientific user facilities generate X-ray scattering data faster than traditional workflows can process them. We address this challenge across two settings, offline dataset exploration and live on-the-fly analysis. We train a domain-specific attent...
170. Some Complexity Results for Robustness Verification for Binarized Neural Networks ​
Author: Harshit Goyal, Sudakshina Dutta
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.CC
arXiv:2606.18918v4 Announce Type: replace Abstract: This paper investigates the computational complexity of verification problems for Binarized Neural Networks (BNNs), in which activations and weights are binary. Specifically, we study three verification problems. First, we prove that checking the s...
171. The Joint Effect of Quantization and Sampling Temperature on LLM Safety Alignment: A Factorial Analysis ​
Author: Hari Prasad, Ritam Pal
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.29581v2 Announce Type: replace Abstract: Modern LLM deployments often combine quantization with higher sampling temperatures to reduce cost, latency, or repetition, yet safety evaluations usually treat these as fixed implementation details. We test whether models that are safe at FP16 wit...
172. Warp RL: Reshaping Base Policy Distributions for Dynamics Adaptation ​
Author: Ethan Hirschowitz, Fabio Ramos
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2606.31043v3 Announce Type: replace Abstract: Residual reinforcement learning adapts a pretrained robot policy by learning an additive correction to its actions. While effective when adaptation amounts to shifting the base policy's action distribution, additive corrections cannot change the di...
173. ECHO: Prune To Act, Trace To Learn With Selective Turn Memory In Agentic RL ​
Author: Zijun Xie, Binbin Zheng, Enlei Gong, Jihua Liu, Yuyang You, Lingfeng Liu, Jiayao Tang, Guanqun Zhao, Aoqi Hu, Zeyu Chen
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.31650v3 Announce Type: replace Abstract: Long-horizon language agents must repeatedly interact with tools, accumulate evidence, and make decisions under bounded context windows. Context-management methods make such rollouts feasible by simplifying past interactions through deletion, foldi...
174. Hierarchical Self-Supervised Representation Learning Framework for Multivariate Time Series Grounded in ECG Analysis ​
Author: Siwon Kim
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2607.01145v2 Announce Type: replace Abstract: Data analysis in the medical domain often encounters scenarios involving a limited target dataset and a large, unannotated dataset with a general distribution. Under such circumstances, self-supervised learning (SSL) methods are highly effective fo...
175. Token Geometry ​
Author: Kathan Shah
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.01455v3 Announce Type: replace Abstract: Language models learn continuous programs over discrete symbols, with the embedding table and LM-head acting as the read/write interface between them. We show that this interface has gradient geometry distinct from dense hidden weights which can be...
176. Asymptotic Preservation and Uniform Accuracy of Diffusion and Flow-Matching Samplers ​
Author: Shiheng Zhang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2607.04113v2 Announce Type: replace Abstract: Diffusion and Gaussian-interpolant flow-matching samplers approach data through a terminal noise floor $\varepsilon$, a singular limit for manifold-supported or rank-deficient data. We study two properties of a complete sampler specification, compr...
177. Operator-on-F complements value-equivalence: a planning-time diagnostic for latent world models ​
Author: Donna Vakalis
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.04464v2 Announce Type: replace Abstract: World-model evaluation for model-based reinforcement learning typically asks whether the learned model predicts reward and value well, which can leave planning-relevant errors in the model's latent rollouts unmeasured. We introduce a complementary ...
178. Efficient Long-Horizon Learning for Learned Optimization ​
Author: Xiaolong Huang, Benjamin Th'erien, James Harrison, Eugene Belilovsky
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.06772v4 Announce Type: replace Abstract: Learned optimization aims to improve upon hand-designed optimizers (e.g., Adam and Muon) by meta-learning small neural network optimizers over a distribution of tasks. While recent work has greatly advanced the architectural design and inductive bi...
179. Fast Rates for Semi-Supervised Learning via Data-Augmentation Graph Regularization ​
Author: Adam M. Oberman
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH
arXiv:2607.07513v2 Announce Type: replace Abstract: Self-supervised learning matches supervised accuracy from a fraction of the labels, but the labeled-sample efficiency behind this has lacked a theoretical explanation. We provide one. Data augmentation induces a similarity graph on the unlabeled da...
180. Avoiding unsafe sets when training with Langevin Dynamics ​
Author: Adam M. Oberman
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, math.AP, stat.ML
arXiv:2607.07538v2 Announce Type: replace Abstract: Training a model with noisy gradient descent can be idealized as overdamped Langevin dynamics, and a natural safety question is to bound the probability $\nu_t(\mathcal{A}_H) = \mathbb{P}(Q_t \in \mathcal{A}_H)$ that the trajectory lies in a design...
181. M+Adam: Low-Precision Training via Additive-Multiplicative Optimization ​
Author: Xiaoyuan Liang, Sebastian Loeschcke, Mads Toftrup, Anima Anandkumar
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.10611v2 Announce Type: replace Abstract: Training with quantized weights can reduce costs but often results in degraded accuracy, especially when optimization is carried out in low precision, without storing high-precision copies. We identify a key failure mode: under low precision, stand...
182. Bandit PCA with Minimax Optimal Regret ​
Author: Mo"ise Blanchard, Dmitrii Ostrovskii, Aadirupa Saha
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.10936v2 Announce Type: replace Abstract: We study the bandit-feedback version of online principal component analysis (Bandit PCA): in each round $t = 1,\dots,T$, the adversary selects a $d \times d$ symmetric gain matrix $G_t$ with spectrum in $[0,1]$ and rank at most $r$; the learner sim...
183. SCOPE-RL: Optimizing Reasoning Paths Before and After Success ​
Author: Xiaojian Liu, Han Xu, Jianqiang Xia, Zhixuan Li, Ke Xu, Yiwei Dai, Xinran Chen, Changwo Wu, Yuchen Li
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.11506v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) optimizes LLMs using sparse verifiable final-answer rewards. This sparse anchor reliably verifies whether a trajectory succeeds but provides no direct feedback on the reasoning path that produce...
184. Generalized Distribution-Free Semi-Supervised Learning with Risk Rewrite ​
Author: Yushi Hirose, Hiroo Irobe, Takafumi Kanamori
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.11947v2 Announce Type: replace Abstract: Typical semi-supervised learning (SSL) methods rely on distributional assumptions, and their performance degrades when these are violated. While PNU learning, a risk rewriting method, offers a distribution-free alternative, it is restricted to bina...
185. When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary ​
Author: Jim Allchin
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.11953v2 Announce Type: replace Abstract: Does a reinforcement-learning agent that earns high reward represent its task's latent state, or only a reward-correlated shortcut? Usually this is unanswerable: the "true state" is undefined. We make it exactly answerable with a white-box instrume...
186. LiteTopK: Exploiting the Curse of Dimensionality for a Fused Indexer-TopK Kernel in Long-Context Sparse Attention ​
Author: Ziqi Yin, Jianyang Gao, Peiqi Yin, Jiangneng Li, Gao Cong
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.11976v2 Announce Type: replace Abstract: Indexer-TopK, the operation to compute the scores and select the top-k candidates, is widely used by sparse attention kernels in large language models and vector retrieval in recommendation systems and vector databases. However, existing GPU-based ...
187. Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs ​
Author: Nikita Kozodoi, Zainab Afolabi, Jack Butler
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.11997v2 Announce Type: replace Abstract: Multi-task model merging combines separately trained expert models into a single model that handles all tasks without co-training. Standard practice merges experts at their optimal validation loss. We challenge this convention by systematically stu...
188. What Does Goodness Measure? A Likelihood-Ratio Account of Forward-Forward Learning ​
Author: Paolo Giannitrapani
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG, eess.IV, stat.ML
arXiv:2607.12501v2 Announce Type: replace Abstract: The Forward-Forward (FF) algorithm trains each layer locally, so that a scalar goodness - the sum of squared activations - is high on real inputs and low on contrastive ones, with activations normalized between layers. Both choices are usually trea...
189. The Spectrum Is Not Enough: When Context Helps Time-Series Forecasting ​
Author: Mert Onur Cakiroglu, Mehmet Dalkilic, Hasan Kurban
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.13006v2 Announce Type: replace Abstract: A growing family of indices scores how predictable a series is from its spectrum. Practitioners increasingly read these scores as answering a different question: whether \emph{adding context}, a longer lookback, a retrieval plug-in, or a pretrained...
190. Fully AI-Generated Image Detection: Definition, Recent Advances and Challenges ​
Author: Qijie Xu, Can Wang, Jiawei Chen, Siwei Lyu, Defang Chen
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2502.19716v3 Announce Type: replace-cross Abstract: Recent advances in visual generative models have enabled the creation of highly realistic, fully AI-generated images without relying on real source content. While beneficial for many applications, these models also pose significant societal r...
191. New universal operator approximation theorem for encoder-decoder architectures ​
Author: Janek G"odeke, Pascal Fernsel
Published: 7/16/2026, 4:00:00 AM
Categories: math.FA, cs.LG, math.GN
arXiv:2503.24092v2 Announce Type: replace-cross Abstract: Motivated by the rapidly growing field of mathematics for operator approximation with neural networks, we present a novel universal operator approximation theorem for broad classes of encoder-decoder architectures and a wide range of input an...
192. Optimizing Binary and Ternary Neural Network Inference on RRAM Crossbars using CIM-Explorer ​
Author: Rebecca Pelke, Jos'e Cubero-Cascante, Nils Bosbach, Niklas Degener, Florian Idrizi, Lennart M. Reimann, Jan Moritz Joseph, Rainer Leupers
Published: 7/16/2026, 4:00:00 AM
Categories: cs.ET, cs.LG
arXiv:2505.14303v3 Announce Type: replace-cross Abstract: Using Resistive Random Access Memory (RRAM) crossbars in Computing-in-Memory (CIM) architectures offers a promising solution to overcome the von Neumann bottleneck. Due to non-idealities like cell variability, RRAM crossbars are often operate...
193. Uniform Approximation of Functions with Asymmetric Growth and Decay by Deep Weighted Polynomials ​
Author: Kingsley Yeon, Steven B. Damelin
Published: 7/16/2026, 4:00:00 AM
Categories: math.NA, cs.AI, cs.LG, cs.NA, stat.ML
arXiv:2506.21306v2 Announce Type: replace-cross Abstract: Functions that grow without bound on one side of the real line and decay to zero on the other cannot be approximated uniformly by ordinary polynomials on unbounded domains. Motivated by classical weighted polynomial approximation, we introduc...
194. Efficiency, Feasibility, and Incentive-Awareness in Constrained Online Resource Allocation ​
Author: Yan Dai, Negin Golrezaei, Patrick Jaillet
Published: 7/16/2026, 4:00:00 AM
Categories: cs.GT, cs.LG, stat.ML
arXiv:2507.09473v2 Announce Type: replace-cross Abstract: We study the dynamic allocation of indivisible resources to strategic agents under long-term constraints, where the planner aims to maximize social welfare, satisfy multiple constraints, and elicit near-truthful reports. We find standard prim...
195. Inverse-LLaVA: Rethinking Multimodal Alignment via Text-to-Vision Mapping ​
Author: Xuhui Zhan, Tyler Derr
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2508.12466v2 Announce Type: replace-cross Abstract: Traditional multimodal learning approaches rely on alignment pre-training to bridge vision and language modalities, typically by projecting visual features into discrete text token spaces using large-scale image--text data. We revisit this de...
196. Deep Unrolling of Sparsity-Induced RDO for 3D Point Cloud Attribute Coding ​
Author: Tam Thuc Do, Philip A. Chou, Gene Cheung
Published: 7/16/2026, 4:00:00 AM
Categories: eess.IV, cs.IT, cs.LG, math.IT
arXiv:2509.08685v2 Announce Type: replace-cross Abstract: Given encoded 3D point cloud geometry available at the decoder, we study the problem of lossy attribute compression in a multi-resolution B-spline projection framework. A target continuous 3D attribute function is first projected onto a seque...
197. Smooth Quasar-Convex Optimization with Constraints ​
Author: David Mart'inez-Rubio
Published: 7/16/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2510.01943v3 Announce Type: replace-cross Abstract: Quasar-convex functions form a broad nonconvex class with applications to linear dynamical systems, generalized linear models, and Riemannian optimization, among others. Current nearly optimal algorithms work only in affine spaces due to the ...
198. BenthiCat: An opti-acoustic dataset for advancing benthic classification and habitat mapping ​
Author: Hayat Rajani, Valerio Franchi, Borja Martinez-Clavel Valles, Raimon Ramos, Rafael Garcia, Nuno Gracias
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2510.04876v3 Announce Type: replace-cross Abstract: Benthic habitat mapping is fundamental for understanding marine ecosystems, guiding conservation efforts, and supporting sustainable resource management. Yet, the scarcity of large, annotated datasets limits the development and benchmarking o...
199. Pretraining in Actor-Critic Reinforcement Learning for Locomotion ​
Author: Jiale Fan, Andrei Cramariuc, Tifanny Portela, Marco Hutter
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2510.12363v4 Announce Type: replace-cross Abstract: The pretraining-finetuning paradigm has facilitated numerous transformative advancements in artificial intelligence research in recent years. However, in the domain of reinforcement learning (RL) for robot locomotion, individual skills are of...
200. Benefits and Limitations of Communication in Multi-Agent Reasoning ​
Author: Michael Rizvi-Martel, Satwik Bhattamishra, Neil Rathi, Guillaume Rabusseau, Michael Hahn
Published: 7/16/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.LG
arXiv:2510.13903v2 Announce Type: replace-cross Abstract: Chain-of-thought prompting has popularized step-by-step reasoning in large language models, yet model performance still degrades as problem complexity and context length grow. By decomposing difficult tasks with long contexts into shorter, ma...
201. Adaptive Conformal Inference through the Lens of Blackwell Approachability ​
Author: Guillaume Principato, Gilles Stoltz
Published: 7/16/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2510.15824v2 Announce Type: replace-cross Abstract: This article considers an online version of conformal inference, called adaptive conformal inference [ACI] and introduced by Gibbs and Cand`es (2021): prediction sets are issued sequentially, after observing features and before the outcomes ...
202. Value Drifts: Tracing Value Alignment During LLM Post-Training ​
Author: Mehar Bhatia, Shravan Nayak, Gaurav Kamath, Marius Mosbach, Karolina Sta'nczak, Vered Shwartz, Siva Reddy
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.CY, cs.LG
arXiv:2510.26707v2 Announce Type: replace-cross Abstract: As LLMs occupy an increasingly important role in society, they are more and more confronted with questions that require them not only to draw on their general knowledge but also to align with certain human value systems. Therefore, studying t...
203. Making Interpretable Discoveries from Unstructured Data: A High-Dimensional Multiple Hypothesis Testing Approach ​
Author: Jacob Carlson
Published: 7/16/2026, 4:00:00 AM
Categories: econ.EM, cs.LG
arXiv:2511.01680v4 Announce Type: replace-cross Abstract: Social scientists are increasingly turning to unstructured datasets to unlock new empirical insights, e.g., estimating descriptive statistics of or causal effects on quantitative measures derived from text, audio, or video data. In many setti...
204. Power Homotopy for Zeroth-Order Non-Convex Optimizations ​
Author: Chen Xu
Published: 7/16/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2511.13592v2 Announce Type: replace-cross Abstract: The existing method of GS-PowerOpt solves the non-convex optimization problem of the form $\max_{\boldsymbol{x} \in \mathbb{R}^d} f(\boldsymbol{x})$ through maximizing a Gaussian-smoothed surrogate $F_{N,\sigma}(\boldsymbol{\mu}) = \mathbb{E}...
205. Leveraging Differentiable PDE Solvers for Semi-Neural Spatial Reconstruction From Sparse Measurements ​
Author: Ofek Aloni, Barak Fishbain
Published: 7/16/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2601.20496v2 Announce Type: replace-cross Abstract: Generating dense physical fields from sparse measurements is a fundamental question in sampling, signal processing, and many other applications. State-of-the-art approaches to this problem either rely on spatial statistics that ignore the gov...
206. Convergence Rates for Distribution Matching with Sliced Optimal Transport ​
Author: Gauthier Thurin (ENS-PSL), Claire Boyer (LMO, IUF, CELESTE), Kimia Nadjahi (ENS-PSL)
Published: 7/16/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2602.10691v2 Announce Type: replace-cross Abstract: We study the slice-matching scheme, an efficient iterative method for distribution matching based on sliced optimal transport. We investigate convergence to the target distribution and derive quantitative non-asymptotic rates. To this end, we...
207. When Audio Separation Hurts Zero-Shot ASR: Evaluating SAM-Audio with Whisper on Bengali and English Speech ​
Author: Akif Islam, Raufun Nahar, Md. Ekramul Hamid
Published: 7/16/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG
arXiv:2603.04710v2 Announce Type: replace-cross Abstract: Recent advances in automatic speech recognition (ASR) and speech enhancement have strengthened the common belief that cleaner audio should lead to more accurate transcription. In this work, we examine whether this assumption holds for modern ...
208. Design-Specification Tiling for ICL-based CAD Code Generation ​
Author: Yali Du, San-Zhuo Xi, Hui Sun, Ming Li
Published: 7/16/2026, 4:00:00 AM
Categories: cs.SE, cs.LG
arXiv:2603.12712v2 Announce Type: replace-cross Abstract: Large language models~(LLMs) have demonstrated remarkable capabilities in code generation, yet their performance remains limited on domain-specific tasks such as Computer-Aided Design~(CAD) code generation, largely due to the scarcity of high...
209. Tactile Modality Fusion for Vision-Language-Action Models ​
Author: Charlotte Morissette, Amin Abyaneh, Wei-Di Chang, Anas Houssaini, David Meger, Hsiu-Chin Lin, Jonathan Tremblay, Gregory Dudek
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG
arXiv:2603.14604v2 Announce Type: replace-cross Abstract: We propose TacFiLM, a lightweight modality-fusion approach that integrates visual-tactile signals into vision-language-action (VLA) models. While advances in VLAs have introduced robot policies that are both generalizable and semantically gro...
210. A plug-and-play approach with fast uncertainty quantification for weak lensing mass mapping ​
Author: Hubert Leterme, Andreas Tersenov, Jalal Fadili, Jean-Luc Starck
Published: 7/16/2026, 4:00:00 AM
Categories: astro-ph.CO, astro-ph.IM, cs.LG, stat.ME
arXiv:2603.22006v2 Announce Type: replace-cross Abstract: Upcoming stage-IV surveys such as Euclid and Rubin will deliver vast amounts of high-precision data, opening new opportunities to constrain cosmological models with unprecedented accuracy. A key step in this process is the reconstruction of t...
211. Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels ​
Author: Yaqi Zhao, Haoliang Sun, Yating Wang, Yongshun Gong, Yilong Yin
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2604.06614v2 Announce Type: replace-cross Abstract: Prompt learning has gained significant attention as a parameter-efficient approach for adapting large pre-trained vision-language models to downstream tasks. However, when only partial labels are available, its performance is often limited by...
212. Robust Explanations for User Trust in Enterprise NLP Systems ​
Author: Guilin Zhang, Kai Zhao, Jeffrey Friedman, Xu Chu, Amine Anoun, Jerry Ting
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2604.12069v3 Announce Type: replace-cross Abstract: Robust explanations are increasingly required for user trust in enterprise NLP, yet pre-deployment validation is difficult in the common case of black-box deployment (API-only access) where representation-based explainers are infeasible and e...
213. A-IC3: Learning-Guided Adaptive Inductive Generalization for Hardware Model Checking ​
Author: Xiaofeng Zhou, Guangyu Hu, Hongce Zhang, Wei Zhang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.LO, cs.LG
arXiv:2604.21688v2 Announce Type: replace-cross Abstract: The IC3 algorithm represents the state-of-the-art (SOTA) hardware model checking technique, owing to its robust performance and scalability. A significant body of research has focused on enhancing the solving efficiency of the IC3 algorithm, ...
214. Discovering Ordinary Differential Equations with LLM-Based Qualitative and Quantitative Evaluation ​
Author: Sum Kyun Song, Bong Gyun Shin, Jae Yong Lee
Published: 7/16/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NE, cs.SC
arXiv:2605.07323v2 Announce Type: replace-cross Abstract: Discovering governing differential equations from observational data is a fundamental challenge in scientific machine learning. Existing symbolic regression approaches rely primarily on quantitative metrics; however, real-world differential e...
215. DIVE: Embedding Compression via Self-Limiting Gradient Updates ​
Author: Dongfang Zhao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.IR, cs.LG
arXiv:2605.20689v2 Announce Type: replace-cross Abstract: High-dimensional language-model embeddings increase storage and search costs, while supervised compressors can overfit when relevance labels are scarce. We present DIVE (Dimensionality reduction with Implicit View Ensembles), a residual compr...
216. The TIME Machine: On The Power of Motion for Efficient Perception ​
Author: Mantas Skackauskas, Xinyue Hao, Laura Sevilla-Lara
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2605.23045v3 Announce Type: replace-cross Abstract: Video representation learning has seen tremendous progress in recent years. This has been driven by many factors, including the scale of training and the success of visual models trained contrastively with language. While these factors have p...
217. When Reasoning Hurts: Source-Aware Evaluation of Frontier LLMs for Clinical SOAP Note Generation ​
Author: Faizan Faisal
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2605.24902v2 Announce Type: replace-cross Abstract: Reasoning-enabled LLMs perform strongly on medical reasoning benchmarks, but it remains unclear whether these gains transfer to structured clinical documentation; we investigate this question using SOAP note generation from clinical dialogue ...
218. Auditing Asset-Specific Preferences in Financial Large Language Models: Evidence from Bitcoin Representations and Portfolio Allocation ​
Author: Wenbin Wu
Published: 7/16/2026, 4:00:00 AM
Categories: q-fin.GN, cs.CY, cs.LG
arXiv:2606.02528v2 Announce Type: replace-cross Abstract: Large language models now power robo-advisors and trading agents, yet whether they carry built-in biases toward specific assets is largely untested. We ask three questions: do LLMs systematically prefer certain financial instruments; can an i...
219. Inverting the Streaming-Diffusion Bottleneck: Video-Rate MLLM-Conditioned Edit Diffusion on a Consumer GPU ​
Author: Yoshiyuki Ootani
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2606.05981v2 Announce Type: replace-cross Abstract: Aggressive distillation of the diffusion U-Net inverts the per-frame bottleneck of real-time text-to-image pipelines: once the denoiser is a 4-step or 1-step distilled student, the text encoder becomes the critical path. This inversion is mos...
220. Fixed-Parameter Tractability of Private Synthetic Data Generation ​
Author: Badih Ghazi, Crist'obal Guzm'an, Pritish Kamath, Alexander Knop, Ravi Kumar, Pasin Manurangsi
Published: 7/16/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, stat.ML
arXiv:2606.11283v2 Announce Type: replace-cross Abstract: We study the problem of generating synthetic data under differential privacy. We establish fixed-parameter tractability (FPT) for this problem where the parameter is the treewidth of the query family's incidence graph. Our algorithms attain o...
221. Learning Red Agent Policy from Observations for Neurosymbolic Autonomous Cyber Agents ​
Author: Ankita Samaddar, Sandeep Neema, Daniel Balasubramanian, Xenofon Koutsoukos
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG, cs.SY, eess.SY
arXiv:2606.18223v2 Announce Type: replace-cross Abstract: With sophisticated cyber-attacks becoming increasingly prevalent, modern networks require intelligent autonomous cyber-defense agents trained via Reinforcement Learning (RL). These agents employ neurosymbolic approaches such as behavior trees...
222. Autonomous Video Generation with Counterfactual Controllability for Self-Evolving World Models ​
Author: Xin Wang, Wenxuan Liu, Tongtong Feng, Wenwu Zhu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2606.24152v2 Announce Type: replace-cross Abstract: Large-scale video generation models are increasingly described as world models because they can learn rich spatiotemporal regularities from visual data. However, we argue that an ideal world model should benefit in a self-evolving generative ...
223. Your Data Manifold is Secretly a Reward Model: Shell-LCC for Text-to-Video Generation ​
Author: Shihao Zhang, Yunzhi Li, Yuguang Yan, Junzhe Zhang, Wei Zhao, Bohan Wang, Hanwang Zhang
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2606.30248v2 Announce Type: replace-cross Abstract: Recent text-to-video (T2V) diffusion models rely heavily on auxiliary reward signals (e.g., via reward models or DPO) to align generated content with human aesthetics and improve realism. These signals, however, incur substantial computationa...
224. Sign in the Air to Unlock: An Interface for authentication in Virtual and Augmented Reality Powered by Point-Voxel Cross-Attention Network ​
Author: Neda Abdolrahimi, Thiru Siddharth, Frank Sicong chen, Vir V Phoha
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.CR, cs.HC, cs.LG
arXiv:2607.01435v2 Announce Type: replace-cross Abstract: Significant advancement of immersive technologies such as Virtual and Augmented Reality (VR/AR) and their integration into diverse aspects of modern life need authentication interfaces that are secure, intuitive, and compatible with embodied ...
225. OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Exploration for ARC-AGI-3 ​
Author: David Courtis, Wenhao Li, Scott Sanner
Published: 7/16/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.01531v2 Announce Type: replace-cross Abstract: Learning how an environment behaves from interaction is central to building agents that adapt to unfamiliar tasks. World models learned with deep networks are flexible but data-hungry and transfer poorly beyond their training distribution. Pr...
226. CaLiSym: Learning Symplectic Dynamics of Real-World Systems through Structured Canonical Lifts ​
Author: Aristotelis Papatheodorou, Pranav Vaidhyanathan, Natalia Ares, Ioannis Havoutis, Gerard J. Milburn
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.06824v2 Announce Type: replace-cross Abstract: Physics-informed learning promises data-efficient and stable dynamics prediction, yet its strongest geometric guarantees have largely remained confined to closed conservative systems. This excludes robotic systems of interest, where actuation...
227. Diagnosing and Calibrating Tool-Call Boundary Drift in Multi-Teacher On-Policy Distillation ​
Author: Jiabin Shen, Guang Chen, Chengjun Mao
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.07050v2 Announce Type: replace-cross Abstract: Agentic language models must learn when to call tools, when to consume tool responses, and when to answer directly. This makes multi-teacher on-policy distillation a natural training strategy: one teacher can specialize in tool calls, another...
228. TheBioCollection: Unified Pre-Training Scale LLM Corpus for Biology ​
Author: Hyunjin Seo, Hyeon Hwang, Gyubok Lee, Jay Shin, Jimin Park, Taesoo Kim, Sanghoon Lee, Hongjoon Ahn, Sungjun Han, Sangwon Jung
Published: 7/16/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG
arXiv:2607.08803v2 Announce Type: replace-cross Abstract: The push toward large language models for biology (BioLM) has created a need for training corpora that can endow models with a genuine understanding of biology. However, existing biological resources, such as molecular databases, protein repo...
229. BiLoG-Net: A Bi-Context Location-Guided Network for Breast Mass Segmentation and Malignancy Classification in Mammography ​
Author: Abu Fatema Mohammad Abdun Noor, Md Imam Ahasan, Md Samiul Ahasan, Kah Ong Michael Goh, S M Hasan Mahmud, Raihana Zannat
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.10188v2 Announce Type: replace-cross Abstract: Breast cancer remains the most commonly diagnosed malignancy among women worldwide, yet accurate detection and characterization of breast masses in mammography remain challenging due to subtle intensity variations, heterogeneous tissue densit...
230. RecRec: Recursive Refinement for Sequential Recommendation ​
Author: Pervez Shaik, Prosenjit Biswas, Abhinav Thorat, Ravi Kolla, Niranjan Pedanekar
Published: 7/16/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.10541v3 Announce Type: replace-cross Abstract: Sequential recommender systems typically infer user preferences through single-pass encoding of interaction histories without iterative refinement, relying on increasingly deep architectures to capture complex patterns. In this work, we revis...
231. Bringing Back Rule Induction to Fluid Intelligence Research? An Initial Validation of the ARC-AGI Benchmark in Humans ​
Author: Jasmin Thelen, Oliver Wilhelm
Published: 7/16/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.11263v3 Announce Type: replace-cross Abstract: Two competing perspectives on fluid intelligence (gf) measures propose that performance is primarily constrained either by working memory capacity or by the ability to induce novel relations. The first perspective is currently dominant in mea...
232. Removable Defects: The Economics and Limits of Deliberate Deficiency ​
Author: Cheng Qian
Published: 7/16/2026, 4:00:00 AM
Categories: econ.EM, cs.AI, cs.LG, stat.ML
arXiv:2607.11983v2 Announce Type: replace-cross Abstract: A specialist tolerates blind spots that a generalist does not. Usually this is treated as a cost to be minimized. We treat it as a design variable: a deficiency can be kept because it pays and removed on demand in the rare situation where it ...
233. VQCSim: When Does Compile-Once Statevector Simulation Beat Generic Quantum Frameworks? ​
Author: Anton Firc, Martin Pere\v{s}'ini, Vojt\v{e}ch Mr'azek, Kamil Malinka, Vojt\v{e}ch Stan\v{e}k, Zbyn\v{e}k Li\v{c}ka, Nouhaila Innan, Walid El Maouaki, Alberto Marchisio, Muhammad Shafique
Published: 7/16/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG
arXiv:2607.11985v2 Announce Type: replace-cross Abstract: Hybrid quantum-classical machine learning workflows repeatedly evaluate many small parametrized circuits during training and model exploration. In this regime, framework dispatch and orchestration overhead often dominate runtime. Prior simula...
234. Analyzing Image Encoder Choices and Graph Homophily in GCN Frameworks for Breast Ultrasound Classification ​
Author: Sabahattin Mert Daloglu, Ceren Coskun, Harvey Castro, Soner Hacihaliloglu, Ilker Hacihaliloglu
Published: 7/16/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2607.12054v2 Announce Type: replace-cross Abstract: Breast ultrasound is widely used for screening, yet automated analysis remains challenging due to speckle noise, acquisition variability, and weak separation of benign and malignant cases in standard ultrasound imaging. Graph convolutional ne...
235. Learning from Complementary Ultrasound Representations for Liver Disease Classification ​
Author: Sabahattin Mert Daloglu, Gokce Bekar, Ceren Coskun, Senanur Sahin, Harvey Castro, Soner Hacihaliloglu, Halley P. Letter, Ilker Hacihaliloglu
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.12062v2 Announce Type: replace-cross Abstract: Differentiating non-alcoholic steatohepatitis (NASH) from non-alcoholic fatty liver disease (NAFLD) using ultrasound remains challenging due to subtle tissue alterations and the limited information available in conventional B-mode imaging. In...
236. Mind the Gap: Promises and Pitfalls of Hierarchical Planning in LeWorldModel ​
Author: Niccol`o Caselli, Francesco Massafra, Samuele Punzo, Salvatore Lo Sardo, Ippokratis Pantelidis, Sathya Kamesh Bhethanabhotla
Published: 7/16/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2607.12547v2 Announce Type: replace-cross Abstract: We investigate whether temporal hierarchy can improve LeWorldModel on long-horizon goal-conditioned control. We introduce Hi-LeWM, an extension that freezes the pretrained low-level LeWM and adds high-level planning over latent subgoals. We e...
237. Label-Decoupled Style Augmentation for Domain Generalization in Multi-Label Remote Sensing Scene Classification ​
Author: Alaa Almouradi, Erchan Aptoula
Published: 7/16/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.12704v2 Announce Type: replace-cross Abstract: Multi-label classification assigns several co-occurring labels to each aerial scene, yet deployed models often encounter data distributions different from their training. Feature-statistics augmentation such as MixStyle, EFDMix, and correlate...