arXiv cs.LG - 2026-07-17 ​
228 items collected.
1. Position: Explainability Research Must Prioritize Foundations over Ad-hoc Methods ​
Author: Michal Moshkovitz, Suraj Srinivas, Lesia Semenova, Nave Frost, Cyrus Rashtchian, Valentyn Boreiko, Shichang Zhang, Himabindu Lakkaraju, Cynthia Rudin, Jennifer Wortman Vaughan
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14123v1 Announce Type: new Abstract: Despite the proliferation of Explainable AI (XAI) techniques -- from feature attributions to sparse autoencoders -- explanations rarely influence real-world workflows. In practice, they are often generated and discarded without guiding meaningful actio...
2. CARPRT: Class-Aware Zero-Shot Prompt Reweighting for Black-Box Vision-Language Models ​
Author: Ruijiang Dong, Zesheng Ye, Jianzhong Qi, Lei Feng, Feng Liu, Gang Niu, Masashi Sugiyama
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.14125v1 Announce Type: new Abstract: Pre-trained vision-language models (VLMs) enable zero-shot image classification by computing the similarity score between an image and textual descriptions, typically formed by inserting a class label (e.g., "cat") into a prompt (e.g., "a photo of a")....
3. Explainable Geospatial AI for Satellite Ground Station Siting Using LiDAR-Derived Terrain Intelligence ​
Author: Shohini Sarkar, Smithi Mahendran, Rishi Chudasama, Varun Mannam, Arav Luthra, Yuvraj Rekhi, Vivek Nadig, Arsh Goenka
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET
arXiv:2607.14127v1 Announce Type: new Abstract: Representative clutter height (RCH) is a key parameter in radio propagation and interference analysis because it captures the dominant height of local obstructions that drive terminal clutter loss. Current practice often relies on fixed clutter heights...
4. Certified Domain Consistency for Multi-Domain Retrieval: Label-Free Per-Domain Contamination Control with Conformal Risk Guarantees ​
Author: Jayakumar Manoharan
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR, cs.SY, eess.SY
arXiv:2607.14157v1 Announce Type: new Abstract: Retrieval over corpora that mix several domains often returns relevant but wrong-domain evidence that ranking metrics miss and that conformal risk control bounds only marginally, under-covering the worst domains. This work introduces C3R, a drop-in con...
5. QFireNet: A Quantum-Enhanced U-Net for Wildfire Segmentation from Sentinel-2 Imagery ​
Author: Jaiman Munshi (IonQ Team, App Dev Club, University of Maryland, College Park), Tanvi Tewary (IonQ Team, App Dev Club, University of Maryland, College Park), Sawyer Bloom (IonQ Team, App Dev Club, University of Maryland, College Park), Aidan Chu (IonQ Team, App Dev Club, University of Maryland, College Park), Chetan Maviti (IonQ Team, App Dev Club, University of Maryland, College Park), Kyon Winston-Bey (IonQ Team, App Dev Club, University of Maryland, College Park), Harshit Badjatia (IonQ Team, App Dev Club, University of Maryland, College Park), Farhan Kittur (IonQ Team, App Dev Club, University of Maryland, College Park), Vardhan Madhavarapu (IonQ Team, App Dev Club, University of Maryland, College Park), Varun Kota (IonQ Team, App Dev Club, University of Maryland, College Park), Joshua Kwon (IonQ Team, App Dev Club, University of Maryland, College Park), Nazia Rangwala-Vohra (IonQ Team, App Dev Club, University of Maryland, College Park), Franz Klein (IonQ Team, App Dev Club, University of Maryland, College Park)
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.14160v1 Announce Type: new Abstract: Wildfire detection from satellite imagery is a semantic image segmentation problem that has proven to be difficult due to challenges such as class imbalance, feature complexity, and atmospheric interference. In this paper, we build on the foundational ...
6. Branching Policy Optimization: Sandbox-Native Language Agent Reinforcement Learning ​
Author: Bowei He, Yankai Chen, Xiaokun Zhang, Xue Liu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.14171v1 Announce Type: new Abstract: Reinforcement learning has emerged as the dominant paradigm for training large language model (LLM) agents that interact with executable sandboxes. State-of-the-art algorithms such as PPO, RLOO, and GRPO inherit their rollout topology from RLHF: for ea...
7. How Much of a 10-K Matters? Aggregation-Dependent Value of Full-Text versus Risk-Factor Sentiment ​
Author: Sanggyu Sean Choi
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP, q-fin.MF, q-fin.ST
arXiv:2607.14174v1 Announce Type: new Abstract: Financial sentiment extraction has largely relied on news text and supervised extraction against return labels alone, leaving 10-K filings -- and volatility, the target risk disclosure is arguably best suited to informing -- comparatively unexplored. W...
8. Low-Latency Relay Selection in NR-V2X Vehicular Communications via Graph Isomorphism Networks with Edge Features ​
Author: Giambattista Amati, Federica Mangiatordi, Emiliano Pallotti, Simone Angelini, Pierpaolo Salvo, Paola Vocca
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.NI
arXiv:2607.14176v1 Announce Type: new Abstract: Reliable, low-latency uplink connectivity is a key requirement for C-V2X networks in dense urban environments, where fast channel variations and blockages often degrade direct vehicle-to-infrastructure links. Multi-hop relaying can restore coverage, bu...
9. RENEW: Towards Learning World Models and Repairing Model Exploitation from Preferences ​
Author: Logan Mondal Bhamidipaty, Mykel Kochenderfer, Subramanian Ramamoorthy
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14180v1 Announce Type: new Abstract: World models are widely used in offline reinforcement learning (RL) to improve sample efficiency and generate experience beyond a fixed dataset. However, they are vulnerable to model exploitation where data coverage is thin. Prior work addresses this e...
10. Closed-Loop Knowledge Dynamics: An Operational Framework for Saturation and Escape ​
Author: Xuening Wu, Shan Yu, Shenqin Yin
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14185v1 Announce Type: new Abstract: Feedback-driven loops support iterative improvement in large language models, reinforcement learning, and autonomous discovery, yet their gains often diminish under repeated internal feedback. We study why closed-loop knowledge systems saturate and wha...
11. A Temporal Machine Learning-Based Time-to-Event Model for Predicting ALS Progression and Healthcare Utilization ​
Author: Zongliang Yue, Qi Li, Terry Heiman-Patterson, Frank Bearoff, Zhaohui Qin, Huanmei Wu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.IR, stat.ML
arXiv:2607.14190v1 Announce Type: new Abstract: Amyotrophic lateral sclerosis (ALS) is a progressive and heterogeneous neurodegenerative disease in which predicting clinically meaningful milestones, such as assistive device use, remains challenging. We developed a time-to-event, digital-twin-inspire...
12. TEDDY: A Pediatric Foundation Model for Risk Forewarning from ICD-Coded Diagnostic Histories ​
Author: Matthew Brady Neeley, Jorge Botas, Johnathan Jia, Lin Yao, Daniel Palacios, Benjamin Choi, Zhandong Liu, Hyun-Hwan Jeong
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14191v1 Announce Type: new Abstract: Pediatric electronic health records capture developmentally structured clinical trajectories, yet their potential for generative healthcare foundation models remains largely unexplored. Here we present TEDDY (Temporal Event Decoder for Disease in Youth...
13. Long-term User Engagement Optimization through Model-agnostic Downstream Rewards Learning ​
Author: Dingsu Wang, Filip Ryzner, Kelly He, Armando Ordorica, David Woo, Aditya Mantha, Liyao Lu, Usha Amrutha Nookala, Haoran Guo, Jiacong He, Olafur Gudmundsson, Matt Chun, Krystal Benitez, Dhruvil Deven Badani, Yijie Dylan Wang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2607.14192v1 Announce Type: new Abstract: As recommender systems mature in the past few years, their optimization objectives have evolved from a primary focusing on short-term behavioral signals to a broader emphasis on long-term user engagement and retention. However, directly optimizing rete...
14. Augmentations for Robust and Efficient Imitation Learning in Streamed Video Games ​
Author: Somjit Nath, Abdelhak Lemkhenter, Pallavi Choudhury, Chris Lovett, Katja Hofmann, Sergio Valcarcel Macua, Lukas Sch"afer
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14200v1 Announce Type: new Abstract: Imitation learning is an appealing way to scale game-playing agents to complex 3D environments by training policies to map visual observations to actions from human demonstrations. However, these demonstrations are expensive to collect and modern game-...
15. Privacy Leakage in Federated Learning in Radiology Reports: A Comparative Evaluation of Tokenizer-Driven Privacy Risks ​
Author: Santhosh Parampottupadam, Andres Martinez, Dimitrios Bounias, Sinem Sav, Klaus Maier-Hein, Ralf Floca
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CR
arXiv:2607.14205v1 Announce Type: new Abstract: Federated learning (FL) enables multi-institutional training on clinical text without sharing raw data, but gradient inversion can reconstruct sensitive information from shared model updates. The extent of this leakage for radiology reports, and the ro...
16. LIGO-PINN: Learned Initialization via Gated Optimization to Alleviate Convergence Failures in Physics Informed Neural Networks ​
Author: Nilay Anurag, Shital Adhikari, Taniya Kapoor, Nikhil Muralidhar
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14233v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have had a broad research impact in modeling domains governed by partial differential equations (PDE). However, PINNs have been shown to perform poorly, sometimes even converging to trivial solutions, in challen...
17. MIDiff: Tackling Sparsity and Imbalance in Mobile Usage Generation via Multivariate-Imaging Diffusion ​
Author: Yilai Liu, Shiyuan Zhang, Hongyang Du
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14249v1 Announce Type: new Abstract: Mobile usage traces are critical for tasks such as user behavior prediction and app recommendation, yet their use is constrained by privacy restrictions and costly large-scale data collection. Although generative models perform well on general time ser...
18. Local Additive Feature Attribution: A Mathematical Taxonomy and Reporting Checklist ​
Author: Rebecca Afriyie Sarpong, Daniel Commey
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14271v1 Announce Type: new Abstract: Feature-attribution methods are central to explainable artificial intelligence. Their assumptions are expressed in several mathematical languages: cooperative-game values, path integrals, gradient operators, perturbation distributions, and backpropagat...
19. Lyapunov Guidance: A Unified Framework for Stabilizing Generative Flows ​
Author: Jingdong Zhang, Xinze Li, Yize Jiang, Luan Yang, Minkai Xu, Junhong Liu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, math.DS, math.OC
arXiv:2607.14272v1 Announce Type: new Abstract: Flow matching has emerged as an effective framework for learning complex data distributions, but adapting pretrained flow models to new tasks often requires computationally expensive retraining. Post-training guidance provides a more efficient alternat...
20. NeuroGRIP: Retrieval-Augmented Graph Refinement for Knowledge-Grounded EEG Seizure Diagnosis ​
Author: Lincan Li, Zheng Chen, Yushun Dong
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14314v1 Announce Type: new Abstract: Seizure diagnosis from EEG signals is a critical yet persistently challenging task, due to the complicated neural dynamics and the spurious connections in inter-channel modeling. While spatial-temporal graph neural networks (STGNNs) have advanced EEG b...
21. Towards a Unified Multidimensional Explainability Metric: Evaluating Trustworthiness in AI Models ​
Author: Georgios Makridis, Georgios Fatouros, Athanasios Kiourtis, Dimitrios Kotios, Vasileios Koukos, Dimosthenis Kyriazis, Jonh Soldatos
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14315v1 Announce Type: new Abstract: In this paper, we present a comprehensive framework for assessing the explainability of various XAI methods, such as LIME and SHAP, across multiple datasets and machine learning models, with the ultimate goal of creating a unified multidimensional expl...
22. Counterfactual Optimal Action Trees (COAT): Interpretable Prescriptive Policies from Observational Data ​
Author: Youssef Drissi, Markus Ettl, Shivaram Subramanian, Wei Sun, Zack Xue
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14318v1 Announce Type: new Abstract: We introduce COAT (Counterfactual Optimal Action Tree), a framework for learning interpretable prescriptive policies from observational data. COAT combines counterfactual outcome estimation with large-scale mixed-integer optimization, using column gene...
23. Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values ​
Author: Jan Betley, Johannes Treutlein, Jan Dubi'nski, Harry Mayne, Karol Ga{\l}\k{a}zka, Niels Warncke, Anna Sztyber-Betley, Owain Evans
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2607.14345v1 Announce Type: new Abstract: People use language models for practical questions whose answers are difficult to verify. We show that models exhibit covert value leakage: the information they provide is influenced by their own values, without this influence being disclosed to the us...
24. Learning Who to Treat When Treatment is Missing ​
Author: Johnna Sundberg, Rayid Ghani, Eli Ben-Michael, Edward Kennedy
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ME
arXiv:2607.14346v1 Announce Type: new Abstract: Policy learning methods are increasingly used to inform treatment allocation under budget constraints. Most proposed methods assume complete treatment data, yet applications frequently suffer from missingness that can bias estimates and lead to subopti...
25. Dysco: Dynamic Subspace Boosting to Mitigate LoRA Interference in Federated Learning ​
Author: Haobo Zhang, Jiankun Wang, Suraj Rajendran, Weishen Pan, Lam Tsoi, Yong Chen, Fei Wang, Jiayu Zhou
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14367v1 Announce Type: new Abstract: Federated fine-tuning of large pre-trained models increasingly relies on Low-Rank Adaptation (LoRA) to reduce communication and computation, but heterogeneous clients can make adapter aggregation unstable. We identify the data-parameter interference as...
26. Supervised Fine-Tuning vs. In-Context Learning: An Equilibrium Analysis of LLM Personalization under Congestion ​
Author: Fengzhuo Zhang, Zhuoran Yang, Dirk Bergemann
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, econ.TH, stat.ML
arXiv:2607.14371v1 Announce Type: new Abstract: Large Language Models (LLMs) have revolutionized AI services, but a critical tension emerges: while personalization improves model performance, it consumes scarce computational resources that users must share. When should a user invest in expensive Sup...
27. A Noise-Robust Elicit-to-Optimize Framework for Distortion Riskmetrics via Inverse Reinforcement Learning ​
Author: Yang Liu, Yuhao Liu, Yunran Wei
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, q-fin.RM
arXiv:2607.14373v1 Announce Type: new Abstract: We propose a noise-robust elicit-to-optimize framework that integrates inverse reinforcement learning (IRL) and reinforcement learning (RL) for eliciting agents' risk preferences and optimizing policies under a broad class of risk objectives characteri...
28. Integration Matters: Rollout-Based Training for Constrained Diffusion Models ​
Author: Xiaoxuan Liang, Saeid Naderiparizi, Berend Zwartsenberg, Frank Wood
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14398v1 Announce Type: new Abstract: Constrained generative models aim to produce samples that satisfy complex feasibility constraints while remaining faithful to the data distribution. Existing constrained generation methods typically enforce constraints either through training-time opti...
29. LATTICE: Graph Self-Supervised Learning for Multimodal Spatial Omics Integration ​
Author: Jagan Mohan Reddy Dwarampudi, Veena Kochat, Suresh Satpati, Kunal Rai, Tania Banerjee
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, q-bio.GN
arXiv:2607.14410v1 Announce Type: new Abstract: Spatially resolved omics studies increasingly combine transcriptomic and epigenomic assays, yet downstream analysis is often still performed using single-modality pipelines. We present LATTICE (Latent Alignment of Tissue-level and Transcriptomic Inform...
30. Adaptive Ad Load Design for Sponsored Search Markets: Evidence, Theory, and Deployment ​
Author: Mohammad Rashid, Hema Yoganarasimhan
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, econ.GN, q-fin.EC
arXiv:2607.14418v1 Announce Type: new Abstract: Ad-load design is a central supply-side decision in sponsored search: more sponsored slots can raise revenue, but may crowd out organic results and degrade user outcomes. We study this trade-off using a large-scale randomized field experiment on an And...
31. HyperShadow: A Benchmark for Detecting 3D Projections of Higher-Dimensional Spatial Objects ​
Author: Akshay Sasi
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CG
arXiv:2607.14419v1 Announce Type: new Abstract: Machine-learning datasets labelled "4D" universally denote three spatial dimensions plus time. We introduce HyperShadow, the first public benchmark in which the fourth, fifth, and sixth dimensions are spatial: the task is to decide whether a 3D point c...
32. Active Real-World Factor-Based Evaluation for Generalist Robot Policies ​
Author: Andrew Liao, Hanchen Cui, Karthik Desingh, Aryan Deshwal
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2607.14439v1 Announce Type: new Abstract: Generalist robot manipulation policies trained on large, diverse datasets have shown remarkable promise across a wide range of tasks. However, rigorously evaluating these policies remains a fundamental challenge. Real-world performance depends on a lar...
33. Depth-Dependent Hidden-State Collapse in Dynamical System Autoencoders for LiDAR Point-Cloud Classification ​
Author: Patricia Medina, Hy P. G. Lam
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.14463v1 Announce Type: new Abstract: We study Dynamical System Autoencoders (DSAE) for LiDAR point-cloud classification using spatial coordinates and Product Coefficient feature augmentations. The experiments compare separately trained DSAE architectures at encoder depths $K=1,\ldots,5$ a...
34. Interleaved Noise Injection Improves Clean, Corrupted, and OOD Performance ​
Author: Matt L. Wiemann, Peter Melchior, Andrew K. Saydjari
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14466v1 Announce Type: new Abstract: Noise injection is a well-known technique in stochastic optimization. We report its surprising effectiveness with an interleaved (on-off-on-off...) rather than the usual monotonic decay schedule. We present a theoretical analysis of noise injection, wh...
35. Non-vacuous Generalization Bounds for Reinforcement Learning with Verifiable Rewards ​
Author: Yuxuan Zhu, Rohan Alur, Daniel Kang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14506v1 Announce Type: new Abstract: While reinforcement learning with verifiable rewards (RLVR) is widely used to improve the reasoning capabilities of large language models (LLMs), the generalizability of the resulting models remains poorly understood. In this work, we establish the fir...
36. Adaptive Runge-Kutta Step Control Buys Training Loss, Not Generalization: An Honest Compute-Matched Study of RK-Adam Optimizers ​
Author: Akhilesh Gogikar
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML
arXiv:2607.14516v1 Announce Type: new Abstract: Interpreting optimizers as gradient-flow discretizations has motivated applying higher-order Runge-Kutta (RK) integrators to neural networks. We build a representative Adam variant (Bogacki-Shampine 3(2) RK pair, FSAL reuse, local-error step control) a...
37. A Continuous-Time Reinforcement Learning Framework for Fine-Tuning Discrete Diffusion Models ​
Author: Zikun Zhang, Jiayuan Sheng, David D. Yao, Wenpin Tang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14522v1 Announce Type: new Abstract: We formulate reinforcement learning (RL) in continuous time with discrete state spaces and possibly arbitrary action spaces via a stochastic control approach, where the state dynamics are modeled as a controlled continuous-time Markov chain (CTMC). We ...
38. xHC: Expanded Hyper-Connections ​
Author: Xiangdong Zhang, Xiaohan Qin, Sunan Zou, Tuo Dai, Xiaoming Shi, Huaijin Wu, Yebin Yang, Zhuo Xia, Shaofeng Zhang, Lin Yao, Yuliang Liu, Yu Cheng, Junchi Yan
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.14530v1 Announce Type: new Abstract: Hyper-Connections (HC) expand the residual stream of Transformers into $N$ parallel streams, providing a form of memory scaling beyond model width and depth. Manifold-Constrained HC (mHC) stabilizes this formulation at scale. The large gains from $N{=}...
39. Muse: Representation Geometry of Muon Beyond Normalized Momentum ​
Author: Da Chang, Qiankun Shi, Lvgang Zhang, Di He, Yaoshuai Ma, Ganzhao Yuan, Yongxiang Liu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14536v1 Announce Type: new Abstract: Muon-style optimizers apply a polar map to matrix momentum, but their updates also depend on the representation of each parameter block before orthogonalization. We study this representation choice as a form of optimizer geometry and introduce {\method...
40. CASP: Learning-Augmented Offline Approximation with Verifiable Certificates and Bounded-Loss PAC Guarantees ​
Author: Haifeng Li, Mo Hai
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14545v1 Announce Type: new Abstract: Machine-learned predictions can speed up offline NP-hard optimization, but asking a predictor what to do amounts to asking it to solve the problem, and committing an unchecked prediction forfeits every worst-case guarantee. CASP (Certificate-Augmented ...
41. Probabilistic Physics-Informed Neural Networks for Estimating Heterogeneous Elastic Properties from Low-Resolution and Noisy Displacement Data ​
Author: Tatthapong Srikitrungruang, Jaesung Lee
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.14563v1 Announce Type: new Abstract: Estimating spatially heterogeneous elastic properties from low-resolution displacement measurements is a severely ill-posed inverse elasticity problem because low resolution obscures spatial details needed to distinguish heterogeneous property variatio...
42. Gate-Zero Growth: A Geometric Framework for Function-Preserving Continual Learning ​
Author: Dante Lok
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14571v1 Announce Type: new Abstract: We introduce \emph{gate-zero growth}, a function-preserving (FP) operator for continual learning that adds new residual blocks through a zero-initialised gate. Under a transversality condition, gate-zero growth induces \emph{rank separation} in the fun...
43. Sharp Stability Threshold and Certification for Designing Stable Residual Architectures ​
Author: Hyemin Gu, Michael Tyrrell, Tuhin Sahai, Markos A. Katsoulakis
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.14576v1 Announce Type: new Abstract: We propose \emph{the sublinear-growth principle} for deep residual architectures -- a sharp stability threshold on the input-magnitude exponent of every residual block's velocity field: $$|v(x, t)| \leq c,|x|^q + b, \qquad q \in [0, 1].$$ The thre...
44. Accelerating A/B-Tests with Counterfactual Estimation: Reducing Variance through Policy Overlap ​
Author: Olivier Jeunen
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.IR, stat.ME
arXiv:2607.14604v1 Announce Type: new Abstract: Online controlled experiments are the gold standard for hypothesis testing in online platforms. Notwithstanding their ubiquity, they are notoriously expensive to run, and issues of variance hamper statistical power in assessing treatment effects. While...
45. Auditing Fairness-Privacy Trade-offs: Subpopulation-Level Effects of Fairness-Enhancing Algorithms ​
Author: Umid Suleymanov, Ilhama Novruzova, Khalid Mammadov, Natavan Hasanova, Murat Kantarcioglu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14607v1 Announce Type: new Abstract: Machine learning (ML) models deployed in sensitive domains such as healthcare, law enforcement, and finance must satisfy not only utility requirements but also fairness and privacy guarantees. While prior work has largely examined how privacy-preservin...
46. Angular Gaussian Supervised Contrastive Learning for Long-Tailed Electrocardiogram Arrhythmia Diagnosis ​
Author: Jin Dai, Qiuzhen Zhang, Chenyun Dai, Danmei Lan, Can Han
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14613v1 Announce Type: new Abstract: Long-tailed label distributions reduce the reliability of deep learning for electrocardiogram (ECG) arrhythmia diagnosis, particularly for clinically important but rare abnormalities. Existing rebalancing and logit adjustment methods mainly address cla...
47. Beyond Entropy: Correctness-Aware Advantage Shaping via Contrastive Policy Optimization ​
Author: Weiwen Xu, Jia Liu, Hou Pong Chan, Long Li, Deng Cai, Min Chen, Hao Zhang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.14614v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) commonly uses entropy for advantage shaping. However, entropy cannot distinguish useful uncertainty from detrimental confusion, limiting its effectiveness as a correctness signal. We propose Contras...
48. PolyQ: Codesigning End-to-End Quantization Framework for Scalable Edge CPU LLM Inference ​
Author: Hyunwoo Oh, Suyeon Jang, Hanning Chen, KyungIn Nam, Sanggeon Yun, Ryozo Masukawa, Mohsen Imani
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.OS
arXiv:2607.14618v1 Announce Type: new Abstract: CPUs are the most universal target for on-device LLM inference, but existing low-bit quantization methods offer either coarse operating points or fine-grained mixed precision that is difficult to execute efficiently on CPUs. We present PolyQ, a CPU-ori...
49. TIDE: Trustworthy and Interpretable Battery Degradation Estimation with Contextual Learning and Symbolic Distillation ​
Author: Wen Yang Tan, Jiawei Li, Fang Liu, Wei Zhang, Sumei Sun, Peng Cheng Wang, Elisa Y. M. Ang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14640v1 Announce Type: new Abstract: Battery health estimation is fundamental for battery management in battery-powered systems, where inaccurate health states may affect control, maintenance, and service life. It becomes even more critical in intelligent connected systems, where estimati...
50. Trajectory-Aware Flow Matching for Topology Optimisation ​
Author: Shusheng Xiao, Jinshuai Bai, Hyogu Jeong, Yunfei Xi, Yilin Gui, YuanTong Gu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2607.14652v1 Announce Type: new Abstract: Topology optimisation (TO) often requires repeated finite element analysis and sensitivity-based material updates, which can be costly when multiple candidate designs are needed under varying physical and design conditions. Generative TO offers a route...
51. Scalable Training of Continuous-Time Spiking Neural Networks with Differentiable Spike-Time Discretization ​
Author: Yusuke Sakemi, Tomoya Takeuchi, Takeo Hosomi, Kazuyuki Aihara
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14672v1 Announce Type: new Abstract: Continuous-time spiking neural networks (SNNs) provide an event-driven framework for temporal computation, computational neuroscience, and neuromorphic hardware. However, training deep continuous-time SNNs is severely constrained by the memory required...
52. Grad2Fair: A Gradient-driven Approach for Graph Fairness without Demographics ​
Author: Yuchang Zhu, Zezhong Xie, Huizhe Zhang, Huazhen Zhong, Jintang Li, Liang Chen, Zibin Zheng
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14705v1 Announce Type: new Abstract: Graph neural networks (GNNs) frequently encounter group fairness issues, often yielding biased predictions against specific demographic groups defined by sensitive attributes such as gender or race. While this challenge has motivated extensive research...
53. MESHA: Mechanism-Enforced Sequential Halving for Strategic Linear Bandits ​
Author: Xin Li, Zixin Zhong
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14706v1 Announce Type: new Abstract: We design and analyze \underline{M}echanism-\underline{E}nforced \underline{S}equential \underline{HA}lving (MESHA), an algorithm for Best Arm Identification (BAI) in strategic linear bandits. In this setting, each arm may strategically misreport its f...
54. Counterfactuals for Feature-Weighted Clustering ​
Author: Richard J. Fawley, Renato Cordeiro de Amorim
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14719v1 Announce Type: new Abstract: Counterfactual explanations provide local, interpretable insight by identifying changes to an input that would alter its assigned outcome. Although well established in supervised learning, their extension to clustering is less direct, since cluster ass...
55. What's in a Smoothness Constant? Tighter Rates for Local SGD with Bounded Second-order Heterogeneity ​
Author: Kumar Kshitij Patel, Rustem Islamov, Sebastian U Stich, Aurelien Lucchi, Eduard Gorbunov, Lingxiao Wang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML
arXiv:2607.14731v1 Announce Type: new Abstract: Local SGD, also known as Federated Averaging, is a widely used distributed optimization algorithm. Although Local SGD often outperforms alternatives such as Mini-batch SGD in practice, theory still only partially explains when and why local updates hel...
56. GAttNHP: Group Attention Neural Hawkes Process for Extrapolation Reasoning in Temporal Knowledge Graphs ​
Author: Xiangni Tian, Kaixian Yu, Runpeng Dai, Niansheng Tang, Hongtu Zhu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.14733v1 Announce Type: new Abstract: Temporal Knowledge Graphs (TKGs) record how facts evolve over time, but forecasting future events on a TKG remains difficult for three reasons: (i) long-range temporal dependencies are hard to encode; (ii) events on different chains mutually excite or ...
57. ChronoQG: Towards a Temporally Expressive and Hop-Bounded Benchmark for Temporal Knowledge Graph Question Generation ​
Author: Xuemeng Liu, Zhengpin Li, Wanpeng Tang, Haotong Xie, Wentao Zhang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14770v1 Announce Type: new Abstract: Knowledge graph question generation (KGQG) aims to generate natural-language questions from structured graph evidence. Existing KGQG benchmarks, however, are mostly built on static knowledge graphs and do not encode the temporal scopes of graph facts. ...
58. Evaluating Epistemic Uncertainty: Beyond OOD Detection and Active Learning ​
Author: Jakub Paplh'am, Willem Waegeman, Eyke H"ullermeier, Vojt\v{e}ch Franc
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14817v1 Announce Type: new Abstract: Current evaluation of epistemic uncertainty relies on tasks such as out-ofdistribution detection and active learning. However, the Bayes-optimal decision strategies for these tasks do not coincide with the scores commonly used to quantify epistemic unc...
59. Asymmetric Peak-Aware Loss for Peak-Critical Time Series Forecasting ​
Author: Theivaprakasham Hari, Yanan Xin, Winnie Daamen, Serge Paul Hoogendoorn, Sascha Hoogendoorn-Lanser
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.14871v1 Announce Type: new Abstract: In many operational time-series forecasting applications, such as crowd demand forecasting, the risk related to under-prediction is substantially higher than that of over-prediction. Accurate prediction of rare demand spikes plays a critical role in do...
60. PAC Learning in Turn-Based Stochastic Games with Reachability Objectives: A Decentralized Private Approach via Expected Conditional Distance ​
Author: Ali Asadi, Krishnendu Chatterjee, Pavol Kebis
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.GT
arXiv:2607.14877v1 Announce Type: new Abstract: Reachability is the most fundamental logical objective, yet it is notoriously difficult to learn in reinforcement learning settings: even for Markov decision processes, PAC learning of reachability is impossible without additional assumptions. This dif...
61. Innocuous-Seeming Data, Latent Ideology: Ideological Generalisation in Finetuned LLMs ​
Author: Robert Graham, Edward Stevinson, Yariv Barsheshat
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CY
arXiv:2607.14888v1 Announce Type: new Abstract: Finetuning language models on small, curated datasets is standard practice for adapting them to specific policies or domains. We show that finetuning on narrow, factually-defensible, moderation-passing data can cause broad ideological shifts across unr...
62. Analytical study of the optimal combination of binary classifiers based on classifiers-induced partitioning of the training set ​
Author: Jean-Marc Brossier, Olivier Lafitte
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.14889v1 Announce Type: new Abstract: This paper studies an optimal linear combination of binary classifiers based on a logical structuration of the dataset via truth tables. The given classifiers partition data into equivalence classes, allowing for a rigorous analysis of the convexified ...
63. Leveraging Instruction Tuning and Merging for Reasoning Model Adaptation ​
Author: Yu-Du Feng, Niels M"undler-Sasahara, Mark Vero, Martin Vechev
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.14895v1 Announce Type: new Abstract: Reasoning language models (RLMs) have demonstrated impressive performance in domains such as mathematics and coding. These domains permit reliable verification of model outputs, which is important for enabling the reinforcement learning that drives RLM...
64. Random Logit Scaling: Defending Deep Neural Networks Against Black-Box Score-Based Adversarial Example Attacks ​
Author: Hamid Dashtbani, Mehdi Dousti Gandomani, AmirMahdi Sadeghzadeh
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2607.14921v1 Announce Type: new Abstract: Machine learning models are increasingly adapted in various domains. However, adversarial examples pose a significant threat to the reliable deployment of these models. In recent years, some powerful adversarial example attacks have been proposed for t...
65. A Minimal Interpretable Architecture for Zero-Shot Reconstruction of Dynamical Systems ​
Author: Christoph J"urgen Hemmer, Florian Plaswig, Daniel Durstewitz
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DS, nlin.CD
arXiv:2607.14937v1 Announce Type: new Abstract: Recent foundation models (FMs) for zero-shot reconstruction of dynamical systems (DS) achieve strong out-of-domain generalization but provide little insight into the mechanisms that underlie their forecasts. Such an understanding could help to strip do...
66. Causal Inference for Sequential Settings under Interference and Latent Confounding ​
Author: Phevos Paschalidis, Constantinos Daskalakis, Devavrat Shah
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, math.PR
arXiv:2607.14940v1 Announce Type: new Abstract: We study causal inference under outcome interference for sequential, observational settings. Specifically, we consider settings where the binary outcomes over N units are Markovian across T time steps. At each time step, the outcomes of N units have de...
67. LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget ​
Author: Changhai Zhou, Kieran Liu, Yuhua Zhou, Qian Qiao, Jun Gao, Harry Zhang, Irvine Lu, Nolan Ho, Lucian Li, Andrew Lei, Cleon Cheng, Steven Chiang, Yihang Zeng, Di Zhang, Rio Yang, Kaijie Chen, Andrew Chen, Pony Ma, Weizhong Zhang, Cheng Jin
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2607.14952v1 Announce Type: new Abstract: A growing gap separates inference context lengths from RL post-training: inference systems are approaching million-token contexts, while post-training workloads often remain at 256K tokens or below and rely on length generalization at deployment. The g...
68. Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation ​
Author: Ku Onoda, Paavo Parmas, Hiroki Furuta, Soichiro Nishimori, Yuta Oshima, Shohei Taniguchi, Yutaka Matsuo
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2607.14962v1 Announce Type: new Abstract: Text-to-image (T2I) models can synthesize realistic, prompt-aligned images, yet samples generated for the same prompt often cover only a small subset of visually distinct modes. This limits the diversity of images, and for person-centric prompts, can r...
69. Multimodal Semantic-Aware Contrastive Learning For False Negative Mitigation in 3D Medical Imaging ​
Author: Sara Ketabi, Matthias W. Wagner, Cynthia Hawkins, Uri Tabori, Birgit Betina Ertl-Wagner, Farzad Khalvati
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2607.14995v1 Announce Type: new Abstract: Multimodal Contrastive Learning (CL) has shown significant performance in aligning representations across various data modalities and improving downstream tasks, especially in healthcare. It works by minimizing the distance between matched (positive) d...
70. Kernel weighted importance sampling for off-policy evaluation in contextual bandits ​
Author: Joshua Spear, Matthieu Komorowski, Rebecca Pope, Neil J Sebire, Erica E. M. Moodie
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.15067v1 Announce Type: new Abstract: This article presents a novel estimator for performing off-policy evaluation using only offline data for contextual bandits. The proposed estimator, Kernel-WIS is demonstrated to be asymptotically consistent and to empirically outperform strong baselin...
71. An Introduction to Sparse Identification of Nonlinear Dynamics for Engineering Applications ​
Author: Yao Cheng Li, Ana Larra~naga, Steven L. Brunton, Urban Fasel
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.15077v1 Announce Type: new Abstract: Many engineering problems involve phenomena whose governing equations are poorly characterized or only partially known. Surrogate modeling techniques such as neural networks can capture the behavior of these systems, but they typically demand large tra...
72. Evaluating covariate balance for long time horizon Markov decision processes ​
Author: Joshua Spear, Rebecca Pope, Neil J Sebire
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.15080v1 Announce Type: new Abstract: This article explores the application of covariate balance diagnostics for detecting the presence of hidden confounding/model miss-specification in studies applying offline reinforcement learning (RL) to deriving optimal treatment recommendations. The ...
73. Learning in Infinitesimal Non-Compositional Sketches ​
Author: Sridhar Mahadevan
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, math.CT
arXiv:2607.15107v1 Announce Type: new Abstract: This paper develops a categorical framework -- Learning in Infinitesimal Non-Compositional Sketches (LINCS) -- as the repair of non-compositionality: failures of diagrams to factor through quotient sketches lifted to the tangent category setting. Machi...
74. On-Policy Delta Distillation ​
Author: Byeongho Heo, Jaehui Hwang, Sangdoo Yun, Dongyoon Han
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.15161v1 Announce Type: new Abstract: On-policy distillation is an alternative post-training method in reinforcement learning that alleviates the constraints imposed by reward models by providing token-level supervision from a teacher model. Although on-policy distillation has been studied...
75. RTS Smoother-Guided Learning of Physics-Based Neural Differential Models ​
Author: Ahmet Demirkaya, Georgios Stratis, Tales Imbiriba, Zachary D. Danziger, Deniz Erdogmus
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2607.15180v1 Announce Type: new Abstract: Ordinary differential equations (ODEs) are widely used to model dynamical systems in physics, biology, neuroscience, and physiology, but in many applications some equations of the dynamics are unknown and only a subset of the state variables are measur...
76. BadWAM: When World-Action Models Dream Right but Act Wrong ​
Author: Qi Li, Xingyi Yang, Xinchao Wang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2607.15207v1 Announce Type: new Abstract: World-action models (WAMs) are emerging as a promising foundation for embodied control: rather than predicting actions alone, they learn representations that couple action generation with future world prediction. This coupling is often viewed as a sour...
77. Data Driven Block Replacement Scheduling ​
Author: Aniruddhan Ganesaraman, VIdyadhar Kulkarni
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.AP, stat.ML
arXiv:2607.15229v1 Announce Type: new Abstract: We develop data-driven algorithms for maintaining $N$ independent identical machines under a \textit{block replacement policy}, in which each machine is replaced upon failure and all machines are jointly replaced at regular intervals of length $k$. The...
78. Mutable Low-Rank Sketches for Retrain-Free Recommendation ​
Author: Hector J. Garcia, Nick Clayton
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.15242v1 Announce Type: new Abstract: A common bottleneck in two-stage recommendation is embedding staleness: when a user rates a new item, their embedding remains fixed until the next retrain cycle. We propose mutable sketches, which store each user's preferences in a KP-tree (a sparse se...
79. Decoding Market Emotion from Blockchain Activity: A Data-Driven Sentiment Classifier ​
Author: Arthur G. Bubolz, Abreu Quevedo, Giancarlo Lucca, Rafael A. Berri, Eduardo Borges, Bruno L. Dalmazo
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2607.15258v1 Announce Type: new Abstract: The growing use of Bitcoin as a decentralized digital asset and investment tool has sparked strong interest in understanding its market behavior. This study presents a new approach to analyze Bitcoin market sentiment by combining on-chain and financial...
80. Generalized Neural Distributional Regression ​
Author: Natan Hilario da Silva, Vicente Garibay Cancho, Adriano Kamimura Suzuki
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.AP, stat.CO, stat.ME, stat.TH
arXiv:2607.14122v1 Announce Type: cross Abstract: We introduce the Generalized Neural Distributional Regression (GNDR) framework, which seamlessly embeds deep neural networks into the parameter space of classical probability distributions. To reconcile the inherent non-identifiability of deep archit...
81. Analysis of Public Schools Educational Performance Based on Causal Models and Hierarchical Clustering ​
Author: Anderson L. de Paula, Pedro C. dos Santos, Renato A. Krohling
Published: 7/17/2026, 4:00:00 AM
Categories: stat.AP, cs.LG
arXiv:2607.14124v1 Announce Type: cross Abstract: The increasing availability of large-scale educational datasets has expanded the use of quantitative methods for investigating school performance. However, institutional heterogeneity among schools and the structural complexity of educational data po...
82. Interpretable Language Model for Closed-Loop Type 1 Diabetes Control ​
Author: Maya Sarkar
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.14126v1 Announce Type: cross Abstract: Type 1 Diabetes (T1D) is a chronic, life-threatening autoimmune condition characterized by the complete destruction of insulin-producing pancreatic beta cells. While Artificial Pancreas Systems (APS) powered by Reinforcement Learning (RL) have shown ...
83. Human AI Construction of Bayesian Networks for Operational Decision Support -- A Virtual Survey Approach ​
Author: Kumar Rahul (Indian Institute of Management Kozhikode, Kerala, India), Shovan Chowdhury (Indian Institute of Management Kozhikode, Kerala, India)
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.14141v1 Announce Type: cross Abstract: Bayesian Belief Networks (BBNs) are powerful tools for decision-making under uncertainty. However, building their structures and estimating parameters are difficult. Currently, researchers must choose between relying on expert judgement or using larg...
84. The Planar Case of Thomas Positive Circuits Conjecture ​
Author: Natan Katz
Published: 7/17/2026, 4:00:00 AM
Categories: math.DS, cs.AI, cs.LG
arXiv:2607.14143v1 Announce Type: cross Abstract: The notion of circuit refers to a cyclic oriented influence between the elements of a dynamical system. There are two classes of circuit: positive and negative. R. Thomas conjectured that a necessary condition of multi stationarity is the existence o...
85. ToolAnchor: Anchoring Counterfactual Context to Boost Agentic Tool-use Capability ​
Author: Weiting Liu, Jieyi Bi, Wanqi Zhou, Jianfeng Feng, Yining Ma, Ai Han, Wenlian Lu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG
arXiv:2607.14145v1 Announce Type: cross Abstract: Tool-augmented large language model agents excel at long-horizon tasks, yet they are typically post-trained on fixed toolsets. When tasks demand new tools, these agents struggle to incorporate them effectively, and retraining from scratch is often im...
86. Breaking Refusal in the First Half: A Mechanistic Study of the Prefill Jailbreak ​
Author: Alex Kwon
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CR, cs.LG
arXiv:2607.14147v1 Announce Type: cross Abstract: Aligned language models refuse harmful requests, but a one-line prefill ("Sure, here is") strips the refusal. We ask where and how it fails. The harm representation stays intact: on the prompts the attack flips to compliance, a linear probe reads har...
87. ITGPT: A Transformer Based Architecture for the Generation of Dance Dance Revolution and In the Groove Charts ​
Author: Miguel O'Malley
Published: 7/17/2026, 4:00:00 AM
Categories: cs.SD, cs.LG
arXiv:2607.14148v1 Announce Type: cross Abstract: Dance Dance Revolution and In the Groove are rhythm games consisting of songs and accompanying choreography, referred to as charts. Players press arrows on a device referred to as a dance pad in time with steps determined by the song's chart. The pro...
88. Enhancing Small Language Models Reasoning through Knowledge Graph Grounding ​
Author: Dimitrios Kelesis, Konstantinos Bougiatiotis, Georgios Paliouras
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.14149v1 Announce Type: cross Abstract: Although large language models (LLMs) have set benchmarks for zero-shot reasoning, their deployment remains cost-prohibitive and environmentally taxing. Small Language Models (SLMs) offer a sustainable alternative, but prone to errors, on tasks requi...
89. Deep-learning Causal Retrieval Optimization for Efficient e-commerce Distribution in Pinterest ​
Author: Junpeng Hou, XianXing Zhang, Sai Xiao, Derek Cheng, Darren Reger, Olafur Gudmundsson, Mehdi Ben Ayed, Zhiqing Rao, Huizhong Duan
Published: 7/17/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2607.14161v1 Announce Type: cross Abstract: Pinterest is where people turn inspiration into action as users browse ideas, then take steps toward realization, often by discovering shoppable content. To support this journey, we must distribute commerce content when it helps, not when it distract...
90. A vision foundation model for single-cell biology via spatial gene cartography ​
Author: Ridvan Yesiloglu, Sakib Mostafa, James Zou, Ash Alizadeh, Jiajun Wu, Lei Xing, Ehsan Adeli, Md Tauhidul Islam
Published: 7/17/2026, 4:00:00 AM
Categories: q-bio.QM, cs.CV, cs.LG
arXiv:2607.14163v1 Announce Type: cross Abstract: Most single-cell foundation models are adapted from language models, representing each cell as a sequence of gene tokens. This discards the relationships among genes and often the magnitude of their expression. We present scVision, a vision foundatio...
91. Towards Reliable AI-Assisted Analog Design: Template-Constrained LLM Agents for SAR ADC Generation ​
Author: Dimple Vijay Kochar, Hae-Seung Lee, Anantha P. Chandrakasan
Published: 7/17/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.AR, cs.LG
arXiv:2607.14165v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have demonstrated significant capability in software code generation, their application to analog Electronic Design Automation (EDA) is bottlenecked. Owing to limited circuit topology understanding and data, directl...
92. When a Verified World Model Still Loses: Play-Adequacy vs Prediction-Accuracy in LLM-Synthesized Code World Models ​
Author: Javier Aguilar Mart'in
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.14169v1 Announce Type: cross Abstract: Large language models can synthesize a game's rules as executable code - a Code World Model (CWM) - which a classical planner then searches over. Such models are typically accepted when they reach high transition accuracy on sampled trajectories. We ...
93. Quantize with Confidence? An Empirical Study of Quantization for Code Generation ​
Author: Saima Afrin, Md. Zahidul Haque, Antonio Mastropaolo
Published: 7/17/2026, 4:00:00 AM
Categories: cs.SE, cs.LG, cs.PL
arXiv:2607.14181v1 Announce Type: cross Abstract: The growing adoption of local inference frameworks such as Ollama has made it increasingly common for developers to run large code models on laptops and other resource-constrained hardware. In these settings, post-training quantization is essential f...
94. NexForge: Scaling Executable Agent Tasks via Requirement-First Synthesis ​
Author: Jiarong Zhao, Zhikai Lei, Zhiheng Xi, Rui Zheng, Hang Yan, Jie Zhou, Qin Chen, Liang He
Published: 7/17/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2607.14186v1 Announce Type: cross Abstract: Scaling executable agent training data is bottlenecked by substrate-first methods that tie task generation to predefined tools, repositories, or skill graphs: expanding coverage requires manual expansion of the substrate, each new domain demands a be...
95. Operator-Informed Gaussian Processes for Complex Helmholtz Wavefields: From Synthetic Benchmarks to In Vivo Brain Elastography ​
Author: Boyuan Deng, Kshitiz Upadhyay, Michael Shields
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.NA, physics.med-ph
arXiv:2607.14193v1 Announce Type: cross Abstract: The Helmholtz equation governs time-harmonic wave propagation, and in dissipative media a complex modulus renders its squared wavenumber $\kappa^2$ complex. Inferring such fields from sparse, noisy data calls for solvers that also quantify their own ...
96. Inference-Time Concept Suppression and Video-Centric Evaluation for Text-to-Video Models ​
Author: Wenxuan Chen, Wenjie Feng
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.14194v1 Announce Type: cross Abstract: Text-to-video (T2V) generators can synthesize realistic and temporally coherent videos, but controllably removing a target concept from a generator remains difficult. Unlike text-to-image concept erasure, T2V unlearning must suppress a target concept...
97. Never Too Late for Force: Accelerating VLA Post-Training with Reactive Force Injection ​
Author: Yi Wang, Wendi Chen, Zimo Wen, Han Xue, Xueqi Li, Wenye Yu, Zhijie Chen, Hao Yang, Jun Lv, Chuan Wen, Cewu Lu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2607.14236v1 Announce Type: cross Abstract: Pretrained vision-language-action (VLA) policies provide strong language-conditioned manipulation knowledge, but they remain largely vision-driven and can struggle once manipulation enters contact states where the scene is occluded, depth is ambiguou...
98. The Steering Budget: Examples beat Knobs ​
Author: Raj Kumar Rajendran
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.14246v1 Announce Type: cross Abstract: Generative models are steered with knobs -- prompts, guidance scales, property tags. Turn one as hard as you like and, past a point, it stops moving the property you care about. We find that ceiling is not a shortcoming of the model but a budget, set...
99. DiMaS: Distribution Matching for Steering Vision-Language-Action Models ​
Author: Pegah Khayatan, Sara Meziane, Jayneel Parekh, Matthieu Cord
Published: 7/17/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.14280v1 Announce Type: cross Abstract: Flow-matching-based vision-language-action (VLA) models have emerged as powerful policies for robotic manipulation, yet a critical capability remains underexplored: fine-grained behavioral control, the ability to govern how a robot performs a task by...
100. XCT-SAM: Sequential Parameter-Efficient Domain Adaptation of SAM for Industrial XCT Defect Segmentation ​
Author: Md Mahedi Hasan, Md Mushfiqur Rahaman, Alan Pachkovskiy, Imtiaz Ahmed, Jeremy Dawson, Srinjoy Das
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.14287v1 Announce Type: cross Abstract: Defect segmentation in additive manufacturing (AM) X-ray computed tomography (XCT) images remains challenging due to severe class imbalance and large distribution shifts across scan conditions. Although recent foundation models such as the Segment An...
101. Real-Time Detection of Charge Jumps in Superconducting Qubits with a Convolutional Neural Network ​
Author: Daniel Gaytan-Villarreal, Peter Meiring, Daniel Baxter, Daniel Bowring, Grace Bratrud, Matteo Cremonesi, Giuseppe Di Guglielmo, Grace Wagner, Bowen Xiao
Published: 7/17/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2607.14293v1 Announce Type: cross Abstract: Ionizing radiation from cosmic rays and gammas can induce discontinuous jumps in the environmental charge of superconducting qubits (charge jumps), causing correlated errors that challenge fault-tolerant quantum computing while simultaneously providi...
102. Spectral Concentration and Recovery in Sparse High-Dimensional Random Geometric Graphs ​
Author: Manuel Fernandez V, Yizhe Zhu
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR, math.ST, stat.TH
arXiv:2607.14304v1 Announce Type: cross Abstract: We study sparse random geometric graphs generated by connecting pairs of high-dimensional vectors whose inner product exceeds a threshold. The latent vectors are sampled either uniformly from the sphere or from a standard Gaussian distribution. Altho...
103. Beyond scalar losses: calibrating segmentation models via gradient vector field surgery ​
Author: Laurin Lux, Alexander H. Berger, Moritz Knolle, Daniel R"uckert, Johannes C. Paetzold
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.14338v1 Announce Type: cross Abstract: Region-based loss functions, such as the Dice loss, have established themselves as the de facto standard for highly class- and region-imbalanced segmentation tasks. However, models trained using region-based loss functions are notoriously miscalibrat...
104. NeuralChaos: Optimal Adapted Approximation of Square Integrable Predictable Processes ​
Author: Anastasis Kratsios, Giulia Livieri, Philipp Schmocker
Published: 7/17/2026, 4:00:00 AM
Categories: math.PR, cs.LG, q-fin.CP, stat.ML
arXiv:2607.14361v1 Announce Type: cross Abstract: We address fundamental challenges in representing and computing $\mathbb{R}^{d}$-valued predictable square-integrable processes over $[0,T]$, collected in the space $\mathcal{H}^2_T(\mathbb{R}^{d})$. These processes are central to continuous-time sto...
105. Random Parameter Noise Does Not Make Exact ReLU Verification Easy ​
Author: Mojtaba Soltanalian
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CC, cs.LG
arXiv:2607.14375v1 Announce Type: cross Abstract: We study exact verification of ReLU networks in an adversarial smoothed model. Every network weight and bias is independently perturbed by Gaussian noise, clipped to $[-2,2]$, and rounded to the exact dyadic grid determined by the input bit complexit...
106. MamaBench: Benchmarking LLM Robustness in Maternal and Child Health Diagnosis through Counterfactual Clinical Perturbation ​
Author: Thanni Adewuyi, Anuoluwa Sotome, Samuel Okoko, Angel Ezendu, Oluwafunke Akinbuwa, Oluwaseun Odunsi, Oluwasegun Oguntuase, Oluwadarasimi Oguntuase, Ifeoma Nwabueze, Abiodun Adereni
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.14385v1 Announce Type: cross Abstract: Large language models achieve strong scores on medical benchmarks, yet these benchmarks evaluate each question in isolation, providing no measure of whether a system can distinguish clinically similar presentations requiring different interventions. ...
107. CatalogAgent: A Supervisor-mediated Self-Learning System Enabling Context Engineering for GenAI Models ​
Author: Zhu Cheng (Xuan), Zhenming Wang (Xuan), Yu (Xuan), Tang, Dan Liu, Bryan Zhang, Athanasios N. Nikolakopoulos, Pranav Souri Itabada, Jing Zhang, Chih-Chi Chou, Peng Gao, Fatemeh Mansoori, Bharat Bojja, Sarath Chander, Sameer Thombare, Umit Batur, Tarik Arici
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.14396v1 Announce Type: cross Abstract: Product catalogs are the backbone of e-commerce sites, yet a large number of structured attributes (SAs) -- such as material, color, and shape -- often have missing values. Typically, SA values are extracted from product information, including titles...
108. Decision Making Needs Uncertainty Quantification [Lecture Notes] ​
Author: Osvaldo Simeone
Published: 7/17/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, math.IT
arXiv:2607.14407v1 Announce Type: cross Abstract: Many signal processing systems ultimately exist to {act}. Whenever the state variable that determines the action to be taken by a decision maker, or agent, is uncertain, the way that uncertainty is represented decides how well the agent performs and ...
109. ConFlow: Constraints-Guided Learning with Flow Matching for Motion Generation ​
Author: Nutan Chen, Jianxiang Feng, Marvin Alles, Botond Cseke
Published: 7/17/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2607.14424v1 Announce Type: cross Abstract: In recent years Flow Matching has become a prominent method for generative modeling robot motion generation. In its generic form Flow Matching is an ODE-based neural sampler that is trained by regressing empirical flow fields associated with motion s...
110. Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel ​
Author: Sietse Schelpe
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.PF
arXiv:2607.14431v1 Announce Type: cross Abstract: We report a way to make a frozen small language model both more capable and dramatically cheaper at once, without changing any weights. Verified knowledge is deposited once as a byte-exact key-value (KV) state artifact and later restored, by graft, i...
111. Can Tokens Compete? Token Representations against Supervised CNN Backbones for BirdCLEF+ 2026 ​
Author: Anthony Miyaguchi, Murilo Gustineli, Adrian Cheung
Published: 7/17/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG
arXiv:2607.14474v1 Announce Type: cross Abstract: This paper details the DS@GT ARC team's approach to BirdCLEF+ 2026, multi-label detection of animal vocalizations in soundscapes from the Pantanal wetlands. The 2026 edition adds about an hour of labeled soundscapes, shifting the task toward supervis...
112. One-Shot Generative Design for Disordered Metamaterials via Self-Organizing Neural Cellular Automata ​
Author: Yujie Xiang, Liwei Wang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CE, cs.LG
arXiv:2607.14475v1 Announce Type: cross Abstract: Disordered metamaterials feature microstructures with inherent randomness and irregularity, enabling them to achieve broader property coverage and superior performance unavailable in their regular counterparts. Despite their promise, designing disord...
113. Full-data accuracy with fewer labels for training and fine-tuning machine-learning force fields ​
Author: Sheng Bi, Yi-Ze Wang, Jun Cheng
Published: 7/17/2026, 4:00:00 AM
Categories: physics.chem-ph, cond-mat.mtrl-sci, cs.LG
arXiv:2607.14486v1 Announce Type: cross Abstract: Machine-learning force fields (MLFFs) are reliable only near their training distribution, making efficient construction of diverse training sets a major bottleneck for both train-from-scratch and foundation fine-tuning workflows. Active learning can ...
114. SAGA: Schema-Aware Grounding for Agentic Text-to-SPARQL Generation ​
Author: Yiming Zhang, Koji Tsuda
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.IR, cs.LG
arXiv:2607.14494v1 Announce Type: cross Abstract: Complex knowledge base question answering (KBQA) is commonly approached through either information retrieval over a question-specific subgraph or semantic parsing into an executable logical form. We study the latter paradigm. Recent large language mo...
115. Multi-Scale ViT Inference with Habitat-Fit Priors and kNN Retrieval for Multi-Species Plant Identification ​
Author: Alper Erten, Murilo Gustineli, Adrian Cheung
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.14509v1 Announce Type: cross Abstract: This paper describes DS@GT ARC's third-place solution to the PlantCLEF 2026 challenge on multi-species plant identification in vegetation quadrat images, where systems must predict every species present in high-resolution (~3000 x 3000 pixel) plot ph...
116. RetroAgent: Harnessing LLMs to Search Over Structured Memory for Agentic Retrosynthesis Planning ​
Author: Yanqiao Zhu, Jingru Gan, Xiaoqi Sun, Fang Sun, Yidan Shi, Md Mofijul Islam, Chao Shang, Wenhao Gao, Connor W. Coley, Yizhou Sun, Wei Wang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.14512v1 Announce Type: cross Abstract: Multi-step retrosynthesis planning seeks to decompose a target molecule into commercially available building blocks through a sequence of feasible reactions. The vast combinatorial search space makes this task challenging even for expert chemists. Tr...
117. Riesz-Kernel Stein Variational Gradient Descent: Renormalized Entropy and Long-Time Particle Limits ​
Author: Trevor Teolis, Maarten V. de Hoop
Published: 7/17/2026, 4:00:00 AM
Categories: math.AP, cs.LG
arXiv:2607.14527v1 Announce Type: cross Abstract: Stein variational gradient descent (SVGD) transports interacting particles toward a target distribution through deterministic kernelized dynamics. Singular Riesz kernels are attractive because they can provide quantitative population-level convergenc...
118. MIDI-RAE-JEPA: Hierarchical Representation Learning and Generation for Symbolic Music ​
Author: Scott H. Hawley
Published: 7/17/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS
arXiv:2607.14537v1 Announce Type: cross Abstract: Rich internal representations of musical structure are essential for music understanding tasks such as machine-assisted music co-writing, yet self-supervised approaches for symbolic music representation remain underexplored, particularly those that e...
119. Advanced Image Generation: Negative Prompt Optimization and Latent Classifier Guidance ​
Author: Vaddi Charan Sai Nandan Reddy, Harini B, Chandana M S
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.14580v1 Announce Type: cross Abstract: We present a novel system that integrates negative prompt optimization via a fine-tuned sequence-to-sequence LLM and latent-space classifier guidance to improve the quality of images generated by Stable Diffusion. Our approach automatically generates...
120. ExaGEMM: Exploration Framework for CPU-Driven ML Inference via Associative In-Register Computing for Low-Bit GEMM ​
Author: Hyunwoo Oh, Suyeon Jang, Hanning Chen, Sanggeon Yun, Ryozo Masukawa, Mohsen Imani
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, cs.OS
arXiv:2607.14622v1 Announce Type: cross Abstract: Low-bit GEMM is increasingly central to efficient ML inference, yet very-low-bit execution remains a poor fit for conventional CPUs. Practical deployment spans fragmented regimes-from 1/2/4-bit weights to varying activation precision-whose feasibilit...
121. Stop Thinking, Start Looking: Efficient Post-Training for Multimodal Document Question Answering via Reasoning-Free Alignment ​
Author: Harikrishnan P M, Goutham Vignesh, Ganesh Parab, Saisubramaniam Gopalakrishnan, Vishal Vaddina, Varun V, Rohit Agrawal
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.14682v1 Announce Type: cross Abstract: Efficient multimodal document question answering with explicit visual grounding, locating the precise document region that supports each answer remains an open challenge. Current approaches bifurcate into Supervised Fine-Tuning (SFT), which requires ...
122. Multimodality as Supervision: Self-Supervised Specialization to the Test Environment via Multimodality ​
Author: Kunal Pratap Singh, Ali Garjani, Rishubh Singh, Muhammad Uzair Khattak, Efe Tarhan, Jason Toskov, Andrei Atanov, O\u{g}uzhan Fatih Kar, Amir Zamir
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.14721v1 Announce Type: cross Abstract: Cross-modal learning, i.e., learning to predict one modality from another, is a fundamental mechanism for self-supervision via leveraging multimodality. Many practical applications, e.g., deploying a household robot, involve devices that are equipped...
123. GeoDetect: Geometric Adversarial Detection for VLPs ​
Author: Afsaneh Hasanebrahimi, Hanxun Huang, Christopher Leckie, James Bailey, Sarah Erfani
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.14737v1 Announce Type: cross Abstract: Vision-language pre-trained models (VLPs) are widely used in real-world applications. However, they remain vulnerable to adversarial attacks. Although adversarial detection methods have demonstrated success in single-modality settings (either vision ...
124. Toward Energy-Efficient and Low-Power Arrhythmia Detection for Wearable Devices ​
Author: Floriaan Bulten, Yawar Rasheed, Arlene John, Vincenzo Stoico, Ghayoor Gillani
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, cs.NE, cs.PF
arXiv:2607.14747v1 Announce Type: cross Abstract: Cardiovascular diseases are the leading cause of death worldwide, and conditions such as arrhythmia often require long-term monitoring for effective detection and diagnosis. However, current wearable monitoring devices are bulky, uncomfortable, and t...
125. Subgrid-Scale Parameterization in Burgers' Equation Using Structure-Preserving Neural Networks and Entropy Variables ​
Author: Aijaz Nazir, Ilya Timofeyev
Published: 7/17/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA
arXiv:2607.14855v1 Announce Type: cross Abstract: We present a machine learning approach for developing subgrid-scale (SGS) parametrizations in coarse simulations of partial differential equations. We utilize structure-preserving neural networks and entropy variables to learn subgrid fluxes in coars...
126. Measuring Spatial Clustering via Metropolis-Hastings Diffusion Distance ​
Author: Thomas Weighill, Chidinma Williams
Published: 7/17/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.TH
arXiv:2607.14880v1 Announce Type: cross Abstract: We propose a novel measure of the discrepancy between two probability distributions $f$ and $g$ on a graph - which we call the diffusion distance - that measures the rate of convergence of $f$ to $g$ under a graph-constrained Markov chain with statio...
127. Domain Adaptation of Mismatched Proximal Denoiser for Plug-and-Play Image Reconstruction ​
Author: Guixian Xu, Jinglai Li, Junqi Tang
Published: 7/17/2026, 4:00:00 AM
Categories: eess.IV, cs.LG, math.OC
arXiv:2607.14894v1 Announce Type: cross Abstract: Plug-and-play proximal gradient descent (PnP-PGD) enables flexible image reconstruction by using denoisers as implicit priors. In practice, these denoisers are often deployed outside their training domains. Existing analyses establish convergence und...
128. Steering Robustness into World Action Models via Mechanistic Interpretability and Optimal Control ​
Author: Jihoon Hong, Julian Skifstad, Qiyue Dai, Alice Chan, Glen Chou
Published: 7/17/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, cs.SY, eess.SY, math.OC
arXiv:2607.14943v1 Announce Type: cross Abstract: World Action Models (WAMs) enable semantically- and physically-informed control but are brittle under distribution shift. In this work, we use mechanistic interpretability to study how robustness-relevant perturbations are represented in WAM activati...
129. Optimal Self-Distillation for Rectified Flow via Linear Probing ​
Author: Saptarshi Roy, Debepsita Mukherjee, Pratik Patil
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.14947v1 Announce Type: cross Abstract: Modern generative models are increasingly trained using model-generated signals, creating both opportunities for self-improvement and risks of collapse. We study optimal self-distillation (SD) for rectified flow (RF): given a suboptimal teacher veloc...
130. SMC-ES: Automated synthesis of formally verified control policies ​
Author: Riccardo Curcio, Toni Mancini, Enrico Tronci
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NE
arXiv:2607.15003v1 Announce Type: cross Abstract: The deployment of autonomous cyber-physical systems in safety-critical environments requires closed-loop control strategies (i.e., policies) that are not only performant but also provably safe and robust. While learning-based methodologies such as Re...
131. cGAP: Generalized Association Plots with HOMALS-Guided Heatmaps for Visualization of High-Dimensional Categorical Data ​
Author: Chun-houh Chen, Shun-Chuan Chang, Chiun-How Kao, Yi-Ju Lee, Shang-Ying Shiu, Yin-Jing Tien, ShengLi Tzeng, Han-Ming Wu
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO, stat.ME
arXiv:2607.15018v1 Announce Type: cross Abstract: High-dimensional categorical data arise in genetics, biomedicine, and the social sciences, yet visualization tools for such data remain far less developed than those for continuous variables. Existing methods either scale poorly, rely heavily on low-...
132. Parameter-efficient Prompt Tuning of Vision Foundation Model With Adaptive Focal Loss for Interpretable MCI Screening ​
Author: Javad Khoramdel, Farhad Hoseyni, Amirhossein Nikoofard
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.15047v1 Announce Type: cross Abstract: Mild Cognitive Impairment is a critical early stage of cognitive decline that frequently precedes Alzheimer's disease, yet its automated detection from neuropsychological drawing tests remains fundamentally constrained by data scarcity, class imbalan...
133. DriftWorld: Fast World Modeling through Drifting ​
Author: Susie Lu, Haonan Chen, Weirui Ye, Yilun Du
Published: 7/17/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG
arXiv:2607.15065v1 Announce Type: cross Abstract: Predictive world models enable robots to plan by imagining the outcomes of their actions, but their value for control hinges on generating many rollouts quickly. This creates a bottleneck for diffusion-based world models: multistep sampling makes eac...
134. AlphaWiSE: Adaptive Weight Interpolation for Continual Multimodal Representation Learning ​
Author: Sarthak Jain, Qiran Hu, Zhen Zhu, Yaoyao Liu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.15094v1 Announce Type: cross Abstract: Multimodal models such as CLIP learn a shared embedding space for cross-modal retrieval, but continual adaptation to sequentially arriving data can disrupt the cross-modal alignment acquired from earlier phases. Conventional continual-learning method...
135. Concept-Guided Spatial Regularization for World Models in Atari Pong ​
Author: Yukuan Lu, Zaishuo Xia, Weyl Lu, Yubei Chen
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.15142v1 Announce Type: cross Abstract: World models are usually evaluated as components of model-based reinforcement learning (MBRL) systems, while the world models themselves are rarely studied in isolation. We examine five representative visual world-model agents in Atari Pong: DreamerV...
136. Subjective Risk Decomposition: A New View for Uncertainty Quantification ​
Author: Raghad Alamri, Michele Caprio, Gavin Brown
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2607.15196v1 Announce Type: cross Abstract: We present a novel viewpoint for uncertainty quantification. Uncertainty measures are not primitives, in need of axioms and argumentation, but instead consequences, of higher-level modelling decisions. We show how epistemic and aleatoric uncertainty ...
137. Mask-Aware Policy Gradients for Diffusion Language Models ​
Author: Haran Raajesh, Kulin Shah, Adam Klivans, Philipp Kr"ahenb"uhl
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.15200v1 Announce Type: cross Abstract: Reinforcement learning has proven effective for improving reasoning in large language models, but extending it to Masked Diffusion Language Models (MDLMs) remains challenging due to the intractability of the log-likelihood estimation. Existing approa...
138. Delocalization of bias in unadjusted Hamiltonian Monte Carlo and underdamped Langevin ​
Author: Yifan Chen, Xiaoou Cheng, Jonathan Niles-Weed, Jonathan Weare
Published: 7/17/2026, 4:00:00 AM
Categories: stat.CO, cs.LG, math.PR, stat.ML
arXiv:2607.15208v1 Announce Type: cross Abstract: Unadjusted samplers such as unadjusted Hamiltonian Monte Carlo and underdamped Langevin are well-known to be biased. Metropolis--Hastings adjustment has been conventionally incorporated into Hamiltonian Monte Carlo to eliminate the bias. However, thi...
139. NeuronSoup: Evolving Asynchronous, Shared-Neuron Temporal Graphs without Backpropagation ​
Author: Subodh Kalia
Published: 7/17/2026, 4:00:00 AM
Categories: cs.NE, cs.LG
arXiv:2607.15217v1 Announce Type: cross Abstract: We present NeuronSoup, a neural computation architecture that replaces synchronous layer-by-layer processing with asynchronous, delay-mediated signal propagation through a pool of shared neurons. Each path in the network routes a continuous-valued si...
140. In-Place Tokenizer Expansion for Pre-trained LLMs ​
Author: Jimmy T. H. Smith, Tarek Dakhran, Alberto Cabrera, Simon S. Lee, Paul Pak, Aditya Tadimeti, Tim Seyde, Maxime Labonne, Alexander Amini, Mathias Lechner
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.15232v1 Announce Type: cross Abstract: A tokenizer fixed at the start of pre-training allocates vocabulary in proportion to the pre-training corpus, reflecting the deployment priorities at that time. When those priorities shift, languages added later are split into many more tokens per wo...
141. Online Neural Space Time Memory for Dynamic Novel View Synthesis ​
Author: Baback Elmieh, Lynn Tsai, Zeman Li, Srinivas Kaza, Tiancheng Sun, Gabor Csapo, Ali Behrouz, Yuan Deng, Stephen Lombardi, Steven M. Seitz, Xuan Luo
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.LG
arXiv:2607.15271v1 Announce Type: cross Abstract: Online novel view synthesis from multi-view streaming videos faces a fundamental trade-off: maintaining a persistent, long-horizon memory to reconstruct temporarily occluded regions while operating under strict real-time constraints. While Test-Time ...
142. MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators ​
Author: Yushi Huang, Xiangxin Zhou, Jun Zhang, Liefeng Bo, Tianyu Pang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.15273v1 Announce Type: cross Abstract: MeanFlow generators achieve fast few-step sampling by predicting average velocities over time intervals, making them attractive for efficient generation. Reinforcement learning (RL) has become a powerful way to align diffusion and flow models with hu...
143. RoboTTT: Context Scaling for Robot Policies ​
Author: Yunfan Jiang, Yevgen Chebotar, Ruijie Zheng, Fengyuan Hu, Yunhao Ge, Jimmy Wu, Tianyuan Dai, Scott Reed, Li Fei-Fei, Yuke Zhu, Linxi "Jim" Fan
Published: 7/17/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2607.15275v1 Announce Type: cross Abstract: Recent robot foundation models operate with single-step or short-history visuomotor context. We introduce Test-Time-Training Robot Policies (RoboTTT), a robot model and training recipe that scale visuomotor context to 8K timesteps, three orders of ma...
144. Bridging the Gap between Newton-Raphson Method and Regularized Policy Iteration ​
Author: Zeyang Li, Chuxiong Hu, Yunan Wang, Guojian Zhan, Jie Li, Yao Lyu, Shengbo Eben Li
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2310.07211v2 Announce Type: replace Abstract: Regularization is a cornerstone of modern reinforcement learning. Regularized policy iteration (RPI) provides a fundamental scheme for solving regularized Markov decision processes (RMDPs), and the widely used soft actor-critic algorithm arises as ...
145. Human-In-The-Loop Machine Learning for Safe and Ethical Autonomous Vehicles: Principles, Challenges, and Opportunities ​
Author: Yousef Emami, Mohammadhossein Homaei, Miguel Guti'errez Gait'an, Luis Almeida, Kai Li, Hui Huang, Zhu Han
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2408.12548v3 Announce Type: replace Abstract: Machine Learning (ML) has become central to Autonomous Vehicles (AVs), supporting perception, prediction, planning, control, and decision-making in dynamic environments. However, achieving full autonomy in cluttered and complex scenarios, such as i...
146. Reinforcement Learning in Switching Non-Stationary Markov Decision Processes: Algorithms and Convergence Analysis ​
Author: Mohsen Amiri, Sindri Magn'usson
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2503.18607v2 Announce Type: replace Abstract: We introduce the Switching Non-Stationary Markov Decision Process (SNS-MDP) framework, in which the environment transitions among a finite set of MDPs governed by a latent Markov chain while the agent observes only the external state. We show that ...
147. Primal-dual algorithm for contextual stochastic combinatorial optimization ​
Author: Louis Bouvier, Thibault Prunet, Vincent Lecl`ere, Axel Parmentier
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2505.04757v2 Announce Type: replace Abstract: This paper introduces a novel approach to contextual stochastic optimization, integrating operations research and machine learning to address decision-making under uncertainty. Traditional methods often fail to leverage contextual information, whic...
148. Generalized Fisher-Weighted SVD: Scalable Kronecker-Factored Fisher Approximation for Compressing Large Language Models ​
Author: Viktoriia Chekalina, Daniil Moskovskiy, Tatiana Matveeva, Andrey Kuznetsov, Evgeny Frolov
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2505.17974v2 Announce Type: replace Abstract: The Fisher information is a fundamental concept for characterizing the sensitivity of parameters in neural networks. However, leveraging the full observed Fisher information is too expensive for large models, so most methods rely on simple diagonal...
149. Fully Offline Reinforcement Learning ​
Author: Mattie Fellows, Clarisse Wibault, Uljad Berdica, Johannes Forkel, Maike Osborne, Jakob N. Foerster
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2505.22442v3 Announce Type: replace Abstract: Offline RL (ORL) promises safe and sample-efficient deployment but existing methods rely on undocumented online interactions for hyperparameter tuning and lack reliable fully offline estimates of initial online performance. We introduce SOReL, a fu...
150. SafeOR-Gym: A Benchmark Suite for Safe Reinforcement Learning Algorithms on Practical Operations Research Problems ​
Author: Asha Ramanujam (Davidson School of Chemical Engineering, Purdue University, West Lafayette, IN), Adam Elyoumi (Davidson School of Chemical Engineering, Purdue University, West Lafayette, IN), Hao Chen (Davidson School of Chemical Engineering, Purdue University, West Lafayette, IN), Sai Madhukiran Kompalli (Davidson School of Chemical Engineering, Purdue University, West Lafayette, IN), Akshdeep Singh Ahluwalia (Davidson School of Chemical Engineering, Purdue University, West Lafayette, IN), Shraman Pal (Davidson School of Chemical Engineering, Purdue University, West Lafayette, IN), Dimitri J. Papageorgiou (Energy Sciences, ExxonMobil Technology and Engineering Company, Annandale, NJ), Can Li (Davidson School of Chemical Engineering, Purdue University, West Lafayette, IN)
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2506.02255v2 Announce Type: replace Abstract: Most existing safe reinforcement learning (RL) benchmarks focus on robotics and control tasks, offering limited relevance to high-stakes domains that involve structured constraints, mixed-integer decisions, and industrial complexity. This gap hinde...
151. The Effect of Stochasticity in Score-Based Diffusion Sampling: a KL Divergence Analysis ​
Author: Bernardo P. Schaeffer, Ricardo M. S. Rosa, Glauco Valle
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2506.11378v3 Announce Type: replace Abstract: Sampling in score-based diffusion models can be performed by solving either a reverse-time stochastic differential equation (SDE) parameterized by an arbitrary stochasticity function or a probability flow ODE, corresponding to setting this stochast...
152. Similarity as Reward Alignment: Robust and Versatile Preference-based Reinforcement Learning ​
Author: Sara Rajaram, R. James Cotton, Fabian H. Sinz
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2506.12529v2 Announce Type: replace Abstract: Preference-based Reinforcement Learning (PbRL) entails a variety of approaches for aligning models with human intent to alleviate the burden of reward engineering. However, most previous PbRL work has not investigated the robustness to labeler erro...
153. Addressing Benchmarking Gaps in Large Language Models for Health and Medicine with Dynamic Red-Teaming ​
Author: Jiazhen Pan (Cherise), Bailiang Jian (Cherise), Paul Hager (Cherise), Yundi Zhang (Cherise), Che Liu (Cherise), Friederike Jungmann (Cherise), Hongwei Bran Li (Cherise), Julian Canisius (Cherise), Chenyu You (Cherise), Junde Wu (Cherise), Jiayuan Zhu (Cherise), Fenglin Liu (Cherise), Yuyuan Liu (Cherise), Niklas Bubeck (Cherise), Moritz Knolle (Cherise), Chen (Cherise), Chen (Cherise), Christian Wachinger, Zhenyu Gong, Cheng Ouyang, Georgios Kaissis, Benedikt Wiestler, Daniel Rueckert
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2508.00923v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to answer health-related questions and support healthcare workflows, yet evidence for their safety still relies heavily on static benchmarks that can rapidly become obsolete or be optimized against...
154. EEG-based AI-BCI Wheelchair Advancement: Transformer-Based Learning with Motor Imagery for Brain Computer Interface ​
Author: Bipul Thapa, Biplov Paneru, Bishwash Paneru, Khem Narayan Poudyal
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC
arXiv:2509.25667v3 Announce Type: replace Abstract: This paper presents an Artificial Intelligence (AI) integrated approach to Brain-Computer Interface (BCI)-based wheelchair development, utilizing a motor imagery right-left-hand movement mechanism for control. The system is designed to simulate whe...
155. What Do Temporal Graph Learning Models Learn? ​
Author: Abigail J. Hayes, Tobias Schumacher, Markus Strohmaier
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2510.09416v4 Announce Type: replace Abstract: Learning on temporal graphs has become a central topic in graph representation learning, with numerous benchmarks indicating the strong performance of state-of-the-art models. However, recent work has raised concerns about the reliability of benchm...
156. Mixtures of SubExperts for Large Language Continual Learning ​
Author: Haeyong Kang, Hee Suk Yoon, Dahua Feng, Chang D. Yoo
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2511.06237v2 Announce Type: replace Abstract: Enabling lifelong learning in LLMs demands resolving the stability-plasticity dilemma (i.e., models must incorporate new knowledge without overwriting prior representations) while maintaining scalability under bounded parameter growth. Existing PEF...
157. A Weak Penalty Neural ODE for Learning Chaotic Dynamics from Noisy Time Series ​
Author: Xuyang Li, John Harlim, Dibyajyoti Chakraborty, Romit Maulik
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, math.DS
arXiv:2511.06609v4 Announce Type: replace Abstract: The accurate forecasting of complex, high-dimensional dynamical systems from observational data is a fundamental task across numerous scientific and engineering disciplines. A significant challenge arises from noise-corrupted measurements, which se...
158. Language as a Wave Phenomenon: Semantic Phase Locking and Interference in Neural Networks ​
Author: Alper Y{\i}ld{\i}r{\i}m, .Ibrahim Y"uceda\u{g}
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2512.01208v5 Announce Type: replace Abstract: In standard Transformer architectures, semantic importance is often conflated with activation magnitude, obscuring the geometric structure of latent representations. To disentangle these factors, we introduce PRISM, a complex-valued architecture de...
159. Energy-Efficient Federated Learning via Adaptive Encoder Freezing for MRI-to-CT Conversion: A Green AI-Guided Research ​
Author: Ciro Benito Raggio, Lucia Migliorelli, Nils Skupien, Mathias Krohmer Zabaleta, Oliver Blanck, Francesco Cicone, Giuseppe Lucio Cascini, Paolo Zaffino, Maria Francesca Spadea
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.DC, physics.med-ph
arXiv:2512.03054v2 Announce Type: replace Abstract: Federated Learning (FL) holds the potential to advance equality in health by enabling diverse institutions to collaboratively train deep learning (DL) models, even with limited data. However, the significant resource requirements of FL often exclud...
160. CluCERT: Certifying LLM Robustness via Clustering-Guided Denoising Smoothing ​
Author: Zixia Wang, Gaojie Jin, Jia Hu, Ronghui Mu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2512.08967v2 Announce Type: replace Abstract: Recent advancements in Large Language Models (LLMs) have led to their widespread adoption in daily applications. Despite their impressive capabilities, they remain vulnerable to adversarial attacks, as even minor meaning-preserving changes such as ...
161. The Challenger: When Do New Data Sources Justify Switching Machine Learning Models? ​
Author: Vassilis Digalakis Jr, Christophe P'erignon, S'ebastien Saurin, Flore Sentenac
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2512.18390v2 Announce Type: replace Abstract: Organizations often have an incumbent predictive model in production when new data sources become available. Because historical training data lack the new features, a challenger model must be trained on a small but growing full-feature dataset. We ...
162. Knowledge-Aware Evolution for Task-Free Streaming Federated Continual Learning with Arbitrary Class Overlap ​
Author: Sixing Tan, Xianmin Liu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2601.19788v2 Announce Type: replace Abstract: Federated Continual Learning (FCL) leverages inter-client collaboration to better balance new knowledge acquisition and old knowledge retention on non-stationary data. However, existing FCL methods struggle to adapt to streaming scenarios where seq...
163. To Grok Grokking: Provable Grokking in Ridge Regression ​
Author: Mingyue Xu, Gal Vardi, Itay Safran
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2601.19791v4 Announce Type: replace Abstract: We study grokking, the onset of generalization long after overfitting, in a classical ridge regression setting. We prove end-to-end grokking results for learning over-parameterized linear regression models using gradient descent with weight decay. ...
164. Selecting Hyperparameters for Tree-Boosting ​
Author: Floris Jan Koster, Fabio Sigrist
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, stat.AP, stat.ML
arXiv:2602.05786v3 Announce Type: replace Abstract: Tree-boosting is a widely used machine learning technique for tabular data. However, its out-of-sample accuracy is critically dependent on multiple hyperparameters. In this article, we empirically compare several popular methods for hyperparameter ...
165. Stabilizing Native Low-Rank LLM Pretraining ​
Author: Paul Janson, Edouard Oyallon, Eugene Belilovsky
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.12429v2 Announce Type: replace Abstract: Foundation models have achieved remarkable success, yet their growing parameter counts pose significant computational and memory challenges. Low-rank factorization offers a promising route to reduce training and inference costs, but the community l...
166. Native Extrapolation Awareness in Flow-Based Conditional Generation ​
Author: Constantinos Tsakonas, Serena Ivaldi, Jean-Baptiste Mouret
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2602.13061v2 Announce Type: replace Abstract: The ability of Flow Matching (FM) to model complex conditional distributions has established it as the state-of-the-art for prediction tasks (e.g., robotics, weather forecasting). However, deployment in safety-critical settings is hindered by a cri...
167. Adaptive Time Series Reasoning via Segment Selection ​
Author: Shvat Messica, Jiawen Zhang, Kevin Li, Theodoros Tsiligkaridis, Marinka Zitnik
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.18645v3 Announce Type: replace Abstract: Time series reasoning tasks often start with a natural language question and require targeted analysis of a time series. Evidence may span the full series or appear in a few short intervals, so the model must decide what to inspect. Most existing a...
168. InfoFlow KV: Information-Flow-Aware KV Recomputation for Long Context ​
Author: Xin Teng, Canyu Zhang, Shaoyi Zheng, Danyang Zhuo, Tianyi Zhou, Shenji Wan
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.05353v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) for long-context question answering is bottlenecked by inference-time prefilling over large retrieved contexts. A common strategy is to precompute key-value (KV) caches for individual documents and selectively r...
169. PhasorFlow: A Python Library for Unit Circle Based Computing ​
Author: Dibakar Sigdel, Namuna Panday
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.15886v4 Announce Type: replace Abstract: We present PhasorFlow, an open-source Python library for computing on the $S^1$ unit circle. Inputs are encoded as complex phasors $z=e^{i\phi}$ on the $N$-torus ($\mathbb{T}^N$); as computation proceeds through unitary wave-interference gates, glo...
170. Contrastive Conformal Sets ​
Author: Yahya Alkhatib, Wee Peng Tay
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2603.26261v2 Announce Type: replace Abstract: Contrastive learning produces coherent semantic feature embeddings by encouraging positive samples to cluster closely while separating negative samples. However, existing contrastive learning methods lack a principled construction of geometric sets...
171. VAN-AD: Visual Masked Autoencoder with Normalizing Flow For Time Series Anomaly Detection ​
Author: PengYu Chen, Shang Wan, Xiaohou Shi, Yuan Chang, Yan Sun, Sajal K. Das
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2603.26842v3 Announce Type: replace Abstract: Time series anomaly detection (TSAD) is essential for maintaining the reliability and security of IoT-enabled service systems. Existing methods require training one specific model for each dataset, which exhibits limited generalization capability a...
172. Synthesizing real-world distributions from high-dimensional Gaussian Noise with Fully Connected Neural Network ​
Author: Joanna Komorniczak
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.09091v2 Announce Type: replace Abstract: The use of synthetic data in machine learning applications and research offers many benefits, including performance improvements through data augmentation and privacy preservation of original samples. This work proposes an efficient synthetic data ...
173. Relational Preference Encoding in Looped Transformer Internal States ​
Author: Jan Kirin
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2604.09870v2 Announce Type: replace Abstract: We investigate how looped transformers encode human preference, training lightweight evaluator heads on frozen Ouro-2.6B loop-iteration states on Anthropic HH-RLHF. v2: an erratum is prepended; the original manuscript is unchanged. A post-publicati...
174. From physical surfaces to human-centric heat stress: LST and UTCI heat mapping reveals nonlinear effects of urban morphology ​
Author: Yuan Wang, Shengao Yi, Xiaojiang Li, Pengyuan Liu, Zhiwei Yang, Ronita Bardhan, Rudi Stouffs
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.22433v2 Announce Type: replace Abstract: Heat exposure connects the built environment and public health, directly shaping the livability and sustainability of urban areas. Understanding the spatial heterogeneity of heat exposure and its drivers is vital for climate-adaptive urban planning...
175. NORACL: Neurogenesis for Oracle-free Resource-Adaptive Continual Learning ​
Author: Karthik Charan Raghunathan, Christian Metzner, Laura Kriener, Melika Payvand
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2604.27031v2 Announce Type: replace Abstract: In a continual learning setting, we require a model to be plastic enough to learn a new task and stable enough to not disturb previously learned capabilities. We argue that this dilemma has an architectural root. A finite network has limited repres...
176. Diagnosing and Mitigating Domain Shift in Permission-Based Android Malware Detection ​
Author: Md Rafid Islam
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.09028v3 Announce Type: replace Abstract: Machine learning-based Android malware detectors often fail in real-world deployment due to domain shift, where models trained on one data source perform poorly on applications from another. This paper presents a comprehensive study on the generali...
177. Models Can Model, But Can't Bind: Structured Grounding in Text-to-Optimization ​
Author: Zhiqi Gao, Albert Ge, Alexander Berenbeim, Nathaniel D. Bastian, Frederic Sala
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.21751v2 Announce Type: replace Abstract: Text-to-optimization requires two separable capabilities: modeling -- choosing the right optimization structure -- and binding -- grounding every coefficient, index, and parameter in the concrete problem data. We study this via Text2Opt-Bench, a sc...
178. Grow-Prune-Freeze Networks: Adaptive & Continual Learning Technique for Olfactory Navigation ​
Author: Kordel K. France, Ovidiu Daescu
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET, cs.RO
arXiv:2605.25170v2 Announce Type: replace Abstract: Training data for olfaction is scattered through disparate, non-standardized datasets that limit the ability to build representative world models. Olfactory navigation is a highly dynamic and non-stationary task that benefits from real-time continu...
179. Orthogonality and Dimensionality in Airline Cluster Analysis using PCA and Kernel PCA ​
Author: Andreas Schlapbach
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, stat.ME
arXiv:2606.08322v2 Announce Type: replace Abstract: This methodological study analyzes the effects of collinearity, effective dimensionality, and cluster stability in a 2023 study of US airline profit cycles from 1995 to 2020 by Renold et al., which uses k-means clustering, principal component analy...
180. Code Correctness Is Linearly Decodable from LLM Hidden States Before Generation ​
Author: Carlo Di Cicco
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.14530v3 Announce Type: replace Abstract: Large language models encode rich information in their hidden states. This work asks whether the correctness of code that Qwen3-4B-Instruct-2507 has not yet generated is already legible in its hidden states, evaluated on a set of 444 tasks from Liv...
181. Some Complexity Results for Robustness Verification for Binarized Neural Networks ​
Author: Harshit Goyal, Sudakshina Dutta
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CC
arXiv:2606.18918v4 Announce Type: replace Abstract: This paper investigates the computational complexity of verification problems for Binarized Neural Networks (BNNs), in which activations and weights are binary. Specifically, we study three verification problems. First, we prove that checking the s...
182. Zero-Shot Quantization for Object Detectors using Off-the-Shelf Generative Models ​
Author: Hyunho Lee, Kyomin Hwang, Hyeonjin Kim, Suyoung Kim, Sunghyun Wee, Nojun Kwak
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.31456v2 Announce Type: replace Abstract: With an increasing number of Object Detection (OD) models being deployed on edge devices, Zero-Shot Quantization for OD (ZSQ-OD) aims to quantize these models when access to the original training data is prohibited. Existing research on Zero-Shot Q...
183. Directional Curvature from Armijo Backtracking: A Low-Cost Sharpness Probe and a Calibration-Free Learning-Rate Safeguard for Adam ​
Author: Ashmitha R, J"org Frochte
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.03998v3 Announce Type: replace Abstract: The local sharpness of the loss, the top Hessian eigenvalue $\lambda_1$, determines the largest stable gradient step, but measuring it normally requires Lanczos or Hessian-vector iterations. We observe that a single Armijo backtracking line search ...
184. Multi-Turn On-Policy Distillation with Prefix Replay ​
Author: Baohao Liao, Hanze Dong, Christof Monz, Xinxing Xu, Li Dong, Furu Wei
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, stat.ML
arXiv:2607.04763v2 Announce Type: replace Abstract: We study on-policy distillation (OPD) for agentic tasks, where an LLM agent interacts with an environment over multiple turns and a student imitates a teacher over these multi-turn interaction histories. Fully online OPD is costly because each upda...
185. The RG-Flow Transformer: Encoding Scale-Free Dynamics in Scarce EEG ​
Author: Dibakar Sigdel
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.11950v2 Announce Type: replace Abstract: Brain field potentials are scale-free: their power spectra follow a $1/f^{\beta}$ law whose aperiodic exponent $\beta$ tracks cortical state, and sleep depth in particular is a shift in $\beta$. We ask whether a transformer endowed with an explicit...
186. Learning Mechanistic Reasoning for Chemical Reactions with Large Language Models ​
Author: Xingyu Dang, Haocheng Tang, Junmei Wang, Yanjun Li
Published: 7/17/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.CL, q-bio.BM
arXiv:2607.12771v2 Announce Type: replace Abstract: Reaction mechanisms consist of the step-by-step sequences of elementary reactions that explain chemical transformations. Learning the mechanism logic is therefore essential for enhancing the fundamental chemical intelligence of large language model...
187. Cross-Cluster Weighted Forests ​
Author: Maya Ramchandran, Rajarshi Mukherjee, Giovanni Parmigiani
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2105.07610v5 Announce Type: replace-cross Abstract: Building trustworthy machine learning algorithms for biological applications requires adapting to data heterogeneity from different sources, batches, distributions, or studies. We propose the 'Cross-Cluster Weighted Forest' (CCWF), an ensembl...
188. A short review on the maximum clique problem algorithms with classical, AI, and quantum methods ​
Author: Raffaele Marino, Lorenzo Buffoni, Bogdan Zavalnij
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cond-mat.dis-nn, cs.DS, cs.LG, math.OC, quant-ph
arXiv:2403.09742v2 Announce Type: replace-cross Abstract: This manuscript provides a comprehensive review of the Maximum Clique Problem, a computational problem that involves finding subsets of vertices in a graph that are all pairwise adjacent to each other. As such, this review is a continuation o...
189. Distributionally Robust Optimization via Iterative Algorithms in Continuous Probability Spaces ​
Author: Linglingzhi Zhu, Yunqin Zhu, Yao Xie
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC
arXiv:2412.20556v2 Announce Type: replace-cross Abstract: We study distributionally robust optimization (DRO) for robust inference when the worst-case distribution is continuous, leading to significant computational challenges due to the infinite-dimensional nature of the optimization problem. Unlik...
190. GenTL: A General Transfer Learning Model for Building Thermal Dynamics ​
Author: Fabian Raisch, Thomas Krug, Christoph Goebel, Benjamin Tischler
Published: 7/17/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2501.13703v2 Announce Type: replace-cross Abstract: Transfer Learning (TL) is an emerging field in modeling building thermal dynamics. This method reduces the data required for a data-driven model of a target building by leveraging knowledge from a source building. Consequently, it enables the...
191. A stochastic smoothing framework for nonconvex-nonconcave minEmax problems with applications to Wasserstein distributionally robust optimization ​
Author: Wei Liu, Muhammad Khan, Gabriel Mancino-Ball, Yangyang Xu
Published: 7/17/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2502.17602v2 Announce Type: replace-cross Abstract: We study a class of stochastic nonsmooth optimization problems in which an outer variable minimizes the expectation of a pointwise maximum. This minimization--expectation--maximization (minEmax) problem arises in Wasserstein distributionally ...
192. Limits of Discrete Energy of Families of Increasing Sets ​
Author: Hari Sarang Nathan
Published: 7/17/2026, 4:00:00 AM
Categories: math.CA, cs.LG, math.MG
arXiv:2504.11302v5 Announce Type: replace-cross Abstract: The Hausdorff dimension of a set can be detected using the Riesz energy. Here, we consider situations where a sequence of points, ${x_n}$, ``fills in'' a set $E \subset \mathbb{R}^d$ in an appropriate sense and investigate the degree to whi...
193. Gibbs randomness-compression proposition ​
Author: M. S"uzen
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2505.23869v4 Announce Type: replace-cross Abstract: A proposition that connects randomness and compression is put forward via Gibbs entropy over set of measurement vectors associated with a compression process. In building this connection, we use a performance of a learning task as a probe of ...
194. A Machine Learning Benchmarking Framework for Lipid Nanoparticle Transfection Efficiency Prediction ​
Author: Asal Mehradfar, Mohammad Shahab Sepehri, Jose Miguel Hernandez-Lobato, Glen S. Kwon, Mahdi Soltanolkotabi, Salman Avestimehr, Morteza Rasoulianboroujeni
Published: 7/17/2026, 4:00:00 AM
Categories: q-bio.QM, cs.CE, cs.LG, q-bio.MN
arXiv:2507.03209v2 Announce Type: replace-cross Abstract: The discovery of new ionizable lipids for efficient lipid nanoparticle (LNP)-mediated RNA delivery remains a major bottleneck in RNA therapeutics development. Recent advances demonstrate the potential of machine learning (ML) models to predic...
195. Modeling and Control of Deep Sign-Definite Dynamics with Application to Hybrid Powertrain Control ​
Author: Teruki Kato, Ryotaro Shima, Kenji Kashima
Published: 7/17/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC
arXiv:2509.19869v2 Announce Type: replace-cross Abstract: Data-driven control increasingly relies on deep models for complex systems whose first-principles models are difficult to obtain. For reliable deployment, however, learned dynamics should respect physical structure and lead to tractable optim...
196. Escaping Model Collapse via Synthetic Data Verification: Near-term Improvements and Long-term Convergence ​
Author: Bingji Yi, Qiyuan Liu, Yuwei Cheng, Haifeng Xu
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2510.16657v3 Announce Type: replace-cross Abstract: Synthetic data has been increasingly used to train frontier generative models. However, recent studies raise key concerns that iteratively retraining a generative model on its self-generated synthetic data may keep deteriorating model perform...
197. Idea2Plan: Exploring AI-Powered Research Planning ​
Author: Jin Huang, Silviu Cucerzan, Sujay Kumar Jauhar, Ryen W. White
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2510.24891v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated significant potential to accelerate scientific discovery as valuable tools for analyzing data, generating hypotheses, and supporting innovative approaches in various scientific fields. In this wo...
198. Skewness-Robust Causal Discovery in Location-Scale Noise Models ​
Author: Daniel Klippert, Alexander Marx
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2511.14441v2 Announce Type: replace-cross Abstract: To distinguish Markov equivalent graphs in causal discovery, it is necessary to restrict the structural causal model. Crucially, we need to be able to distinguish cause $X$ from effect $Y$ in bivariate models, that is, distinguish the two gra...
199. DualHNIE: Dual-Channel Hypergraph Learning for Node Importance Estimation in Heterogeneous Knowledge Graphs ​
Author: Jiawen Chen, Yanyan He, Qi Shao, Mengli Wei, Duxin Chen, Wenwu Yu, Yanlong Zhao
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2512.12477v3 Announce Type: replace-cross Abstract: Estimating node importance in heterogeneous knowledge graphs is a fundamental problem underlying recommendation, search, and knowledge decision systems. However, most existing methods rely on pairwise message passing mechanisms that fail to c...
200. ROOFS: RObust biOmarker Feature Selection ​
Author: Anastasiia Bakhmach, Paul Dufoss'e, Simon Charpigny, Florence Monville, Laurent Greillier, Fabrice Barl'esi, S'ebastien Benzekry
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2601.05151v3 Announce Type: replace-cross Abstract: Feature selection (FS) is essential for biomarker discovery and clinical predictive modeling. Over the past decades, methodological literature on FS has become rich and mature, offering a wide spectrum of algorithmic approaches. However, much...
201. Neural Architectures for Amortized Bayesian Inference: Statistical Foundations and Empirical Assessments ​
Author: Roy Shivam Ram Shreshtth, Arnab Hazra, Gourab Mukherjee
Published: 7/17/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO
arXiv:2601.07944v2 Announce Type: replace-cross Abstract: Since the turn of the century, approximate Bayesian inference has steadily evolved as new computational techniques have been incorporated to handle increasingly complex, large-scale predictive problems. The recent success of deep neural netwo...
202. Allure of Craquelure: A Variational-Generative Approach to Crack Detection in Paintings ​
Author: Laura Paul, Holger Rauhut, Martin Burger, Samira Kabri, Tim Roith
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.NA, math.NA
arXiv:2602.09730v3 Announce Type: replace-cross Abstract: Recent advances in imaging technologies, deep learning and numerical performance have enabled non-invasive detailed analysis of artworks, supporting their documentation and conservation. In particular, automated detection of craquelure in dig...
203. Human Motion Data Alone Does Not Guarantee Plausible Gait Biomechanics ​
Author: Xinyi Liu, Jangwhan Ahn, Edgar Lobaton, Jennie Si, He Huang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2603.12408v2 Announce Type: replace-cross Abstract: Motion imitation learning (IL) is increasingly used in robotics and human gait modeling, yet its ability to recover biomechanically consistent joint moments without explicit kinetic information remains unclear. In this study, we examined whet...
204. AgentWorm: Self-Propagating Attacks Across LLM Agent Ecosystems ​
Author: Yihao Zhang, Zeming Wei, Xiaokun Luan, Chengcan Wu, Zhixin Zhang, Jiangrong Wu, Haolin Wu, Huanran Chen, Jun Sun, Meng Sun
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG, cs.MA, cs.SE
arXiv:2603.15727v3 Announce Type: replace-cross Abstract: Autonomous LLM-based agents increasingly operate as long-running processes forming densely interconnected multi-agent ecosystems, whose security properties remain largely unexplored. Systems such as OpenClaw, an open-source platform with over...
205. Automated identification of Ichneumonoidea wasps via YOLO-based deep learning: Integrating HiresCam for Explainable AI ​
Author: Joao Manoel Herrera Pinheiro, Gabriela Do Nascimento Herrera, Alvaro Doria Dos Santos, Luciana Bueno Dos Reis Fernandes, Ricardo V. Godoy, Eduardo A. B. Almeida, Helena Carolina Onody, Marcelo Andrade Da Costa Vieira, Angelica Maria Penteado-Dias, Marcelo Becker
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2603.16351v2 Announce Type: replace-cross Abstract: Accurate taxonomic identification of parasitoid wasps within the superfamily Ichneumonoidea is essential for biodiversity assessment, ecological monitoring, and biological control programs. However, morphological similarity, small body size, ...
206. Autoregressive Guidance of Deep Spatially Selective Filters using Bayesian Tracking for Efficient Extraction of Moving Speakers ​
Author: Jakob Kienegger, Timo Gerkmann
Published: 7/17/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, cs.SD
arXiv:2603.23723v2 Announce Type: replace-cross Abstract: Deep spatially selective filters achieve high-quality enhancement with real-time capable architectures for stationary speakers of known directions. To retain this level of performance in dynamic scenarios where only the speakers' initial dire...
207. Unsupervised Evaluation of Deep Audio Embeddings for Music Structure Analysis ​
Author: Axel Marmoret
Published: 7/17/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG
arXiv:2603.27218v2 Announce Type: replace-cross Abstract: Music Structure Analysis (MSA) aims to uncover the high-level organization of musical pieces. State-of-the-art methods are often based on supervised deep learning, but these methods are bottlenecked by the need for heavily annotated data and ...
208. Photonic convolutional neural network with pre-trained in situ training ​
Author: Saurabh Ranjan, Sonika Thakral, Amit Sehgal
Published: 7/17/2026, 4:00:00 AM
Categories: cs.ET, cs.LG, physics.optics
arXiv:2604.02429v2 Announce Type: replace-cross Abstract: Convolutional neural networks (CNNs) have transformed image processing, but the energy consumption and inference latency of electronic based implementations remain fundamental bottlenecks. These limitations have motivated the search for alter...
209. Towards Predicting Multi-Vulnerability Attack Chains in Software Supply Chains from Software Bill of Materials Graphs ​
Author: Laura Baird, Armin Moin
Published: 7/17/2026, 4:00:00 AM
Categories: cs.SE, cs.CR, cs.LG
arXiv:2604.04977v2 Announce Type: replace-cross Abstract: Software supply chain security compromises often stem from cascaded interactions of vulnerabilities, for example, between multiple vulnerable components. Yet, Software Bill of Materials (SBOM)-based pipelines for security analysis typically t...
210. Lecture notes on Machine Learning applications for global fits ​
Author: Jorge Alda
Published: 7/17/2026, 4:00:00 AM
Categories: hep-ph, cs.LG
arXiv:2604.07520v2 Announce Type: replace-cross Abstract: These lecture notes provide a comprehensive framework for performing global statistical fits in high-energy physics using modern Machine Learning (ML) surrogates. We begin by reviewing the statistical foundations of model building, including ...
211. An architectural capacity ceiling, not a barren plateau: why a fixed-encoding variational quantum circuit cannot fit the Lorenz-63 attractor ​
Author: Tushar Pandey
Published: 7/17/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2604.23743v2 Announce Type: replace-cross Abstract: Variational quantum circuits train poorly on chaotic forecasting, usually blamed on barren plateaus (exponentially vanishing gradients). Using an exactly simulable four-qubit variational quantum physics-informed circuit fit to Lorenz-63, we s...
212. New non-Euclidean neural quantum states from additional types of hyperbolic recurrent neural networks ​
Author: H. L. Dao
Published: 7/17/2026, 4:00:00 AM
Categories: quant-ph, cond-mat.dis-nn, cs.LG
arXiv:2604.24337v2 Announce Type: replace-cross Abstract: In this work, we extend the class of previously introduced non-Euclidean neural quantum states (NQS) which consists only of Poincare hyperbolic GRU, to new variants including Poincare RNN as well as Lorentz RNN and Lorentz GRU. In addition to...
213. FlowBot: Inducing LLM Workflows with Bilevel Optimization and Textual Gradients ​
Author: Hongyeon Yu, Young-Bum Kim, Yoon Kim
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2604.26258v3 Announce Type: replace-cross Abstract: LLM workflows, which coordinate structured calls to individual LLMs/agents to achieve a particular goal, offer a promising path towards building powerful AI systems that can tackle diverse tasks. However, existing approaches for building such...
214. Agentic Vulnerability Reasoning on COTS Binaries ​
Author: Hwiwon Lee, Jongseong Kim, Lingming Zhang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2605.05000v2 Announce Type: replace-cross Abstract: LLM agents have been increasingly adopted for solving security tasks. However, existing evaluations usually require source code access, while commercial off-the-shelf (COTS) binaries dominate deployed software and require reasoning from strip...
215. MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems ​
Author: Xinle Deng, Ruobin Zhong, Hujin Peng, Xiaoben Lu, Yanzhe Wu, Guang Li, Buqiang Xu, Yunzhi Yao, Jizhan Fang, Haoliang Cao, Junjie Guo, Yuan Yuan, Ziqing Ma, Yuanqiang Yu, Rui Hu, Baohua Dong, Hangcheng Zhu, Ningyu Zhang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2605.28732v3 Announce Type: replace-cross Abstract: Memory is essential for enabling large language models to support long-horizon reasoning, yet existing memory systems remain unreliable and difficult to debug. Tracing memory's dynamic evolution is crucial to understand how information is syn...
216. Representing Research Attention as Contextually Structured Flows ​
Author: Jessica Rodrigues, Angelo Salatino, Gard Jenset, Scott Hale
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2606.05895v4 Announce Type: replace-cross Abstract: Research metrics use attention as evidence of societal impact. Yet attention serves as evidence only once interpreted, and its meaning depends on its contextual structure, not on volume alone. Altmetrics represents signals in isolation, keepi...
217. DNQ: Deep Nash Q-Network for Partially Observable n-Player Games ​
Author: Qintong Xie, Edward Koh, Xavier Cadet, Peter Chin
Published: 7/17/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2606.06480v2 Announce Type: replace-cross Abstract: Many real-world competitive systems require multiple decision-makers to act simultaneously under shared constraints, limited information, and repeated interaction, as in auctions, resource allocation, and security competition. We study multi-...
218. Conan-embedding-v3: Fusing Modality-Specific Models for Omni-Modal Embedding ​
Author: Shiyu Li, Zhiyuan Hu, Yifan Wang, Peiming Li, Zheng Wei, Yang Tang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.MM, cs.AI, cs.LG
arXiv:2606.09331v2 Announce Type: replace-cross Abstract: Omni-modal retrieval promises a single embedding space for text, image, video, document, and audio inputs, but building such a unified retriever is difficult since these modalities differ in data distribution, architecture, and optimization d...
219. CANN-EUCLID: unsupervised constitutive artificial neural network model discovery from full-field data ​
Author: Benjamin Alheit, Siddhant Kumar, Mathias Peirlinck
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CE, cs.LG, physics.comp-ph
arXiv:2606.14565v2 Announce Type: replace-cross Abstract: Constitutive artificial neural networks (CANNs) provide interpretable material model discovery, but have so far been used in stress-supervised settings based on apparent stress-strain data from homogeneous tests. Because each test samples onl...
220. When Does Belief-Based Agent Memory Help? Reliability-Conditional Updating and Provenance-Capped Poisoning Defense ​
Author: Pranav Singh
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.IR, cs.LG
arXiv:2606.22030v2 Announce Type: replace-cross Abstract: We investigate when belief-based memory actually improves large language model (LLM) agents. Our vehicle is Nous, a long-term memory architecture that represents each entity-attribute pair as a categorical probability distribution updated thr...
221. EvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures ​
Author: Bu\u{g}ra Alperen Ulu{\i}rmak, Rifat Kurban
Published: 7/17/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.SE
arXiv:2606.30219v2 Announce Type: replace-cross Abstract: LLM evaluation and AI safety face a shared measurement problem: benchmark scores, reward-model signals, and reported safety metrics can improve while the latent properties they are meant to represent remain difficult to verify. This paper com...
222. ADP: Adversarial Dynamics Priors for Physically Grounded Humanoid Locomotion ​
Author: Seokju Lee, Jeongtae Lee, Jeonghyeok Lim, Jeonguk Kang, Byungwook Lee, Seungho Han, Keun Ha Choi, Dongil Park, Kyung-Soo Kim
Published: 7/17/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.03454v2 Announce Type: replace-cross Abstract: In this paper, we propose Adversarial Dynamics Priors (ADP) for perturbation-resilient humanoid locomotion control. Existing motion prior-based methods induce natural motion styles by imitating kinematic motion features, but they do not direc...
223. Governing Generative AI Across Financial Institutions: A Framework for Generative AI Risk Control ​
Author: Dennis Mao, Alessandra Lin, Yixin Kang, Yiqing Wang
Published: 7/17/2026, 4:00:00 AM
Categories: q-fin.RM, cs.LG
arXiv:2607.04103v3 Announce Type: replace-cross Abstract: Generative artificial intelligence is moving from general-purpose experimentation toward specialized applications across banking, capital markets, insurance, payments, and wealth management. Its main contribution is not limited to conversatio...
224. Is the Geometry Doing the Work? An Operating-Point Audit of Hierarchy in Hyperbolic Vision-Language Models ​
Author: Jaeyoung Kim, Eunseok Kim, Dongsuk Jang
Published: 7/17/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.05268v3 Announce Type: replace-cross Abstract: Hyperbolic vision-language models are designed to encode abstraction geometrically: general concepts near the origin, specific ones farther out, and entailment cones representing directed order. We ask whether trained MERU, HyCoCLIP, and PHyC...
225. SKooP: Symmetric Koopman Predictions for Faster and More Generalizable Legged Robot Locomotion with Reinforcement Learning ​
Author: Evelyn D'Elia, Weishu Zhan, Giulio Turrisi, Giulio Romualdi, Giuseppe L'Erario, Raffaello Camoriano, Wei Pan, Daniele Pucci
Published: 7/17/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.11624v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) algorithms classically suffer from poor sample efficiency. In robotics, a recent line of work has emerged addressing this problem by encoding physics priors in the learning process. However, most of these approache...
226. FlashDiff: Efficient Regional Execution and Scheduling for Diffusion Model Serving ​
Author: Yaqi Qiao, Ping He, Songrun Xie, Ayush Barik, Chensong Zhang, Zhengzhong Tu, Fan Lai
Published: 7/17/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2607.12121v2 Announce Type: replace-cross Abstract: Diffusion models have become the central backbone for modern image, video, and audio generation, but their efficient service remains a challenge. Unlike autoregressive decoding, diffusion inference repeatedly updates high-dimensional spatial ...
227. When Directional Accuracy Lies: A Base-Rate-Honest Benchmark for LoRA-Adapted TimesFM on Equity Forecasting ​
Author: Taizhen Cheung
Published: 7/17/2026, 4:00:00 AM
Categories: q-fin.ST, cs.LG
arXiv:2607.12248v2 Announce Type: replace-cross Abstract: Large pretrained time-series models such as TimesFM are attractive for financial forecasting, but raw directional accuracy is a misleading scoreboard in equity markets. An early LoRA adapter in this project appeared to reach roughly 80% direc...
228. Early Adoption of Agentic Coding Tools by GitHub Projects ​
Author: Maliha Noushin Raida, Daqing Hou
Published: 7/17/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CY, cs.LG
arXiv:2607.14037v2 Announce Type: replace-cross Abstract: Agentic coding tools are increasingly capable of generating and submitting pull requests (PRs) to software projects, introducing new forms of human-agent collaboration in software development. While prior studies have examined PR-level outcom...