arXiv cs.LG - 2026-08-20 ​
237 items collected.
1. Proactive Road Safety Intervention in Australia: Predicting Risky Driving Hotspots from Connected Vehicle Data ​
Author: Adriana-Simona Mih\u{a}i\c{t}\u{a}, Clarence Cheung, Artur Grigorev, Tuo Mao, David Lillo-Trynes
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CY
arXiv:2608.16913v1 Announce Type: new Abstract: Road safety monitoring has historically been reactive, relying on crash-record analysis after fatalities and injuries have already occurred. Proactive identification of high-risk locations and dangerous driving behaviour before incidents occur is a cri...
2. Detecting and Discriminating Operator Misspecification in Hybrid PDE-Parameter Learning: a Reference-Free Instrument, with Discrimination Bounded In Sample ​
Author: Eric Fock
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, stat.ME
arXiv:2608.16925v1 Announce Type: new Abstract: We build an instrument that reads, from a single fit and with no oracle, whether the operator a hybrid PDE-parameter estimator postulates is wrong-and separates that from a merely unidentifiable parameter. On one self-adjoint parabolic inverse problem,...
3. Data-DPO: Direct Preference Optimization for Target Model Data Selection in LLM Post-Training ​
Author: Peng Sun, Yi Yang, Antong Zhang, Chunxiao Li, Yanbo Wang, Dianbo Liu, xin chen, Kai Yu, Lu Chen, Tianfan Fu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16926v1 Announce Type: new Abstract: Data selection in supervised fine-tuning aims to select a small set of effective samples from large-scale candidate data, reducing training cost while preserving model performance. However, existing methods usually treat data value as a relatively stat...
4. Hierarchical Data Selection via Manifold Coverage and Sparse Feature Coverage in LLM Post-training ​
Author: Peng Sun, Yi Yang, Antong Zhang, Chunxiao Li, Yanbo Wang, Dianbo Liu, xin chen, Kai Yu, Lu Chen, Tianfan Fu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.16927v1 Announce Type: new Abstract: As supervised fine-tuning data continues to scale, selecting high-value subsets from large candidate pools is crucial for reducing training cost and improving model performance. Existing methods often measure diversity directly in the original embeddin...
5. Benchmarking Classical and Transformer-Based Models for Document Sensitivity Classification ​
Author: Aleesha Zainab, Muhammad Ahmed Khalid, Faheem Ullah Khan, Asifullah Khan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16928v1 Announce Type: new Abstract: Automatic sensitivity classification of organizational documents is a critical yet underserved problem, where the consequences of misclassification range from regulatory violations to security breaches. While AI-based approaches offer a scalable altern...
6. Mr.Dec: Daily-Scale Longitudinal Multimodal Modeling for 30-Day Readmission Prediction ​
Author: Minjun Kim, Jong Hak Moon
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16929v1 Announce Type: new Abstract: Predicting 30-day hospital readmission is essential for assessing patient stability and optimizing healthcare resources. As clinical risk evolves with the accumulation of evidence during hospitalization, capturing these dynamic trajectories is essentia...
7. EMAN: Optimization-Driven Capacity Growth through Path Emergence in Multi-Task Learning ​
Author: Chenlei Fang, Jingchen Li, Hongzong LI, Qingyao Li, Yixuan Zhang, Huarui Wu, Haobin Shi, Chunjiang Zhao
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16930v1 Announce Type: new Abstract: Existing multi-task learning methods rely on hard sharing, multiple paths or experts, adaptive sharing, and dynamic expansion. However, their capacity changes are usually constrained by predefined structures or triggered by task boundaries and conflict...
8. SW-ProxyCE: Zero-Query Adversarial Transfer from Public EEG Encoders to Private Downstream Models ​
Author: Linhua Cong, Dingkun Liu, Dongrui Wu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16931v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models have recently emerged as a promising paradigm for EEG decoding by learning reusable representations from large-scale heterogeneous neural recordings. However, the open release of EEG foundation encoders, w...
9. DOW-KE: Anchor-Free Multi-Layer Knowledge Editing via Direct End-to-End Weight Optimization ​
Author: Ran Chen, Junbo Zhang, Qianli Zhou, Xinyang Deng, Wen Jiang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16932v1 Announce Type: new Abstract: Multi-layer locate-then-edit methods for knowledge editing first optimize target residual-stream activations (anchors) at selected layers, then realize them layer by layer as weight updates. This pipeline optimizes an intermediate representation but de...
10. Study-Strategy Clusters from EdNet Logs Track Engagement, Not Mastery ​
Author: Qingchuan Lyu, Yingxin Li, Albert Yang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CY, stat.AP
arXiv:2608.16963v1 Announce Type: new Abstract: Learning analytics often treats unsupervised clusters of intelligent tutoring system (ITS) logs as learner types that should predict learning. We test that assumption on EdNet-KT3. Clustering study-strategy features (resource use, revision, video, prob...
11. RoBell-RVFL: A Robust Generalized Bell Random Vector Functional Link Network ​
Author: A. Rahaman, A. Quadir, M. Tanveer
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16965v1 Announce Type: new Abstract: The dominance of majority classes in real-world datasets poses a fundamental challenge to randomized neural networks, often biasing decision boundaries and overlooking critical minority samples. Existing remedies, such as synthetic minority over-sampli...
12. MultiSigBERT: Beyond Survival Analysis through Multimodal and Sequential Modeling in Oncology ​
Author: Paul Minchella, St'ephane Chr'etien, Guillaume Metzler, Lo"ic Verlingue, R'emi Vaucher
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16972v1 Announce Type: new Abstract: Machine learning has become an essential component of modern healthcare, where the integration of heterogeneous data sources offers unprecedented opportunities to improve clinical decision-making. Electronic Health Records (EHR) contain complementary i...
13. Position: Fairness Failure in Generative Models is an Evaluation Problem ​
Author: Mariia Vladimirova, Jean-Yves Franceschi, Thibaut Issenhuth
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16974v1 Announce Type: new Abstract: Despite groundbreaking advancements in generative models during the last decade, concerns about their lack of fairness, reinforcing societal inequalities and harming marginalized groups, remain under-addressed and difficult to act upon. This position p...
14. Agents unlock new capabilities through Switching LoRA Adapters as a Tool (SLAaaT) ​
Author: Kenneth Ge
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17034v1 Announce Type: new Abstract: Post-training can unlock new capabilities and improve performance on specialized tasks, but sometimes at the cost of catastrophic forgetting in other domains. This poses a problem in long agent trajectories that compose different capabilities. We rejec...
15. J-Miner: Recovering Executable Decision Knowledge from Language-Model Classifiers ​
Author: Yunfan Gao, Xinyi Huang, Tao Sheng, Haorui Song, Yun Xiong, Haofen Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.17063v1 Announce Type: new Abstract: Large language models can be fine-tuned into specialized classifiers that perform well across diverse text tasks and make complex judgments, but they typically expose only final labels, leaving the decision knowledge acquired through fine-tuning implic...
16. Certified but Private: Scalable Zero-Knowledge Proofs for Neural Network Guarantees ​
Author: Youwei Zhong, Ben Merbaum, Timos Antonopoulos, Ning Luo, Charalampos Papamanthou, Katerina Sotiraki, Ruzica Piskac
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.LO
arXiv:2608.17070v1 Announce Type: new Abstract: With the growing deployment of machine learning models, formal guarantees of the robustness and fairness of these models have become increasingly important in safety-critical and legal-compliance settings. However, model parameters are often commercial...
17. Dynamic Regime-Aware Conformal Calibration for Reliable Economic Forecast Intervals under Multiple Distribution Shifts ​
Author: Bogdan Oancea
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17079v1 Announce Type: new Abstract: Conformal prediction provides distribution-free prediction intervals but relies on exchangeability, an assumption often violated in economic forecasting because of covariate shift, concept drift, local heterogeneity and latent regimes. We propose Dynam...
18. Backward through Time, Algebraically ​
Author: Konstantinos Kogkalidis
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.LO, cs.PL, cs.SY, eess.SY
arXiv:2608.17087v1 Announce Type: new Abstract: Linear temporal logic is a modal extension of propositional logic that allows one to state how a system should behave over time. Its canonical domain is the booleans, but discretely-valued judgements are of little use in steering softly-valued systems ...
19. Deep Learning for Cross-Border Electricity Price Forecasting: A Comparative Study ​
Author: Hadeer Elashhab, Sai Srijan Papineni, Marvin Dorn, Veit Hagenmeyer, Benjamin Sch"afer
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17091v1 Announce Type: new Abstract: While publicly available electricity market data presents a valuable resource for forecasting research, the field lacks established benchmark datasets for standardized comparison. As a result, many studies have relied on different datasets and metrics ...
20. From Abductive Explanations to Global Logical Rules for Node Classification in SGCs ​
Author: Bryan Lima Cavalcante, Thiago Alves Rocha
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.LO
arXiv:2608.17103v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have achieved remarkable performance in node classification tasks, motivating growing interest in methods capable of explaining their predictions. Recent logic-based approaches, such as LogicXGNN, derive global logical rule...
21. Causal Discovery in Equal Variance Linear Gaussian DAGs via SURE-Tuned Ridge Regression ​
Author: Sambit Mishra, Urbashi Mitra
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, stat.ML
arXiv:2608.17132v1 Announce Type: new Abstract: Recovering the directed acyclic graph (DAG) of a structural equation model (SEM) from observational data is a central problem in causal discovery. The iterative gradient descent and per-problem hyperparameter tuning of continuous-optimization methods a...
22. Iterative tensor network transformations for element-wise evaluation of elementary and filtering functions ​
Author: Xiao Wang, Tomohiro Hashizume, Pia Siegl, Dieter Jaksch
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, cs.AI, physics.comp-ph, quant-ph
arXiv:2608.17135v1 Announce Type: new Abstract: Tensor networks are powerful formats for compressing large-scale data. However, their application to general data processing has been limited by the difficulty of performing nonlinear operations. Here, we introduce iterative tensor network transformati...
23. OraclePhys: A Systematic Framework for LLM Fine-Tuning on Structural Mechanics ​
Author: Mingyu Li, Guorui Song, Jing Lin, Haoqian Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17162v1 Announce Type: new Abstract: What a language model internalizes from fine-tuning is usually diagnosed after the fact. We make it an experimental variable. OraclePhys is a systematic fine-tuning framework with three components: OraclePhys-Bench, an exactly-graded structural-mechani...
24. Q-Learning With World Models ​
Author: Perry Dong, Yueru Jia, Chelsea Finn, Dorsa Sadigh
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17163v1 Announce Type: new Abstract: Off-policy reinforcement learning (RL) has become increasingly sample-efficient, enabling applications such as RL fine-tuning of Vision-Language-Action models into reliable, high-performing policies. World models offer a further lever for sample effici...
25. SCENARIODIFF: A Scenario-level Guidance Framework for Multimodal Time Series Forecasting--Extended Version ​
Author: Tuan-Binh Tran, Dat Nguyen Cong, Duc-Trong Le, Thanh Trung Huynh, Tung Kieu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17164v1 Announce Type: new Abstract: Textual context such as news, reports, and logs can provide valuable signals for time series forecasting, especially when future dynamics are driven by external events that are not yet visible in historical values. Existing multimodal forecasting metho...
26. Population Health-Based Machine Learning Reveals Associations Between Psychosocial Factors and Chronic Kidney Disease ​
Author: Md. Atik Shams, David Eisenberg, Sumaiya Fatema, Asma Sultana, D. M Hasibul Islam, Junnatul Mawa, Anindita Datta, Nafiya Ahmed, Danastan Tasaouf Mridula, SK. Sazid Mahmud, Simon Bin Akter, Tanjila Helaly, Jorge Fresneda Fernandez, Humayera Islam, Tanmoy Sarkar Pias
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17174v1 Announce Type: new Abstract: Chronic kidney disease (CKD) progresses silently and severely undermines quality of life, making early detection critical for improving patient outcomes. We present a two-part study that combines large-scale telehealth data with advanced machine learni...
27. Task Specialization Fine-Tuning for Contextual Reinforcement Learning ​
Author: Jianan Zhou, Jung-Hoon Cho, Tianyue Zhou, Han Zheng, Jie Zhang, Roy Dong, Yining Ma, Cathy Wu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17180v1 Announce Type: new Abstract: Contextual Reinforcement Learning (CRL) seeks to generalize classical RL by maximizing task coverage across a context space of related tasks. While prior works often train from scratch and rely on either multi-task learning for a single policy or strat...
28. Reinforcement Learning as (Discrete) Potential Theory ​
Author: Christopher Connolly
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.GT
arXiv:2608.17181v1 Announce Type: new Abstract: Reinforcement learning (RL) theory fundamentally depends on probability theory through the Markov chain. There is a deep connection between probability theory and potential theory. This paper reviews that connection and explores the potential-theoretic...
29. How smoothing the affinity matrix affects neighborhood preservation in t-SNE ​
Author: Shirin Mohebi, Guillaume Bied, Jefrey Lijffijt
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.17190v1 Announce Type: new Abstract: Dimensionality reduction methods are instrumental to visualize high-dimensional data, and t-SNE stands as one of the most widely used methods due to its emphasis on local neighborhood preservation. A central component of t-SNE is the affinity matrix, w...
30. Pessimistic Meta-Induction and Its Limits: Lessons from Frequentist Statistics and Machine Learning Theory ​
Author: Hanti Lin
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ME
arXiv:2608.17213v1 Announce Type: new Abstract: This paper challenges the pessimistic meta-inductive argument against scientific realism by undermining its inductive step rather than its historical premise. Although related challenges already exist, I develop a new one. Drawing on a general epistemo...
31. Delta2Gamma: Band-Wise Adaptive Contrastive Learning of EEG for Alzheimer's Disease Detection ​
Author: Chanwoo Park, Chanwoo Kim
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17231v1 Announce Type: new Abstract: Low-cost, scalable screening for dementia remains an open problem. Imaging-based diagnosis is costly and hard to deploy widely. Electroencephalography (EEG) is portable and inexpensive, but its recordings are noisy, vary widely across subjects, and car...
32. Physics-Informed and Hybrid Machine Learning in Additive Manufacturing: Application to Fused Filament Fabrication ​
Author: Berkcan Kapusuzoglu, Sankaran Mahadevan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, stat.CO
arXiv:2608.17246v1 Announce Type: new Abstract: This article investigates several physics-informed and hybrid machine learning strategies that incorporate physics knowledge in experimental data-driven deep-learning models for predicting the bond quality and porosity of fused filament fabrication (FF...
33. Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL ​
Author: Yunhao Yang, Yuexin Bian, Yunjie Tian, Di Fu, Tianjin Huang, Yuanyuan Shi, Ziang Xiao, Nuno Vasconcelos, Yijiang Li
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.17253v2 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a powerful approach for improving reasoning in language and vision-language models, yet its strongest successes still depend heavily on ground-truth supervision (e.g., verifiable reward). Such annotations are ...
34. Understanding Curriculum Learning in Large Language Models via Cross-Difficulty Optimization Dynamics ​
Author: Zhikai Ding, Ziyi Ye
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17268v1 Announce Type: new Abstract: Curriculum learning has been widely adopted in the post-training of large language models by organizing training data from easy to hard. However, its effectiveness varies substantially across reasoning tasks, suggesting that no single curriculum is uni...
35. Rethinking Irregular Time Series Forecasting from the Perspective of Basis Functions ​
Author: Rongwen Li, Changjian Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17284v1 Announce Type: new Abstract: Irregular time series forecasting is crucial in many domains, such as healthcare and meteorological observation. However, due to the inherent characteristics of irregular time series, including sparse observations and non-uniform sampling, accurately p...
36. Abra: Scaling Diffusion Image Training ​
Author: Kyle Chickering, Wei-An Lin, Swayam Bhanded, Dan Saunders, Akshat Tripathi, Jiaming Song, Shyamal Buch, Xinchen Yan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17286v1 Announce Type: new Abstract: Compute-optimal scaling laws guide the training of frontier language models yet remain largely unexplored for visual generation. We present a systematic scaling law study for text-to-image diffusion models using Abra, a controlled family of flow-matchi...
37. Beyond MSE: Rethinking the Evaluation Metric and Benchmarking for Irregular Time Series Forecasting ​
Author: Rongwen Li, Haixin Xie, Xiao Wang, Changjian Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17293v1 Announce Type: new Abstract: Existing research on irregular time-series forecasting has primarily focused on model design, while evaluation metrics remain insufficiently studied. Existing benchmarks typically use mean squared error (MSE) as the evaluation metric. We show that, in ...
38. Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements ​
Author: Zhi Zheng, Rongsheng Chen, Yunpeng Ba, Zhenkun Wang, Yee Whye Teh, Wee Sun Lee
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17310v1 Announce Type: new Abstract: Reinforcement Learning (RL) has been promising in single-turn LLM fine-tuning. However, long-horizon agentic reasoning introduces increasingly branching interactions and sparse rewards, exposing several limitations of RL: its heavyweight backpropagatio...
39. MoFE: A Novel Mixture-of-Experts Framework with Fourier Neural Operators for Cryptocurrency Forecasting ​
Author: Bowen Liu, Mingming Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17342v1 Announce Type: new Abstract: Forecasting cryptocurrency prices remains a formidable challenge due to inherent non-stationarity, abrupt regime shifts, and multi-scale stochastic dependencies. Conventional deep learning models often struggle to capture complex underlying dynamics, f...
40. Tight Bounds for Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function ​
Author: Anh Tuan Nguyen, Viet Anh Nguyen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.17343v1 Announce Type: new Abstract: Data-driven algorithm design frames hyperparameter tuning as a statistical learning problem, but establishing generalization guarantees remains challenging due to the implicit, non-smooth dependence of model performance on hyperparameters. Existing mul...
41. Repetition as Reinforcement: Enhancing Sample Efficiency via Instant Episode Repetition in Reinforcement Learning ​
Author: Hoda Yamani, Yuning Xing, Koen van Rijnsoever, Bruce A. MacDonald, Henry Williams
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2608.17347v1 Announce Type: new Abstract: Repetition is a fundamental mechanism in human learning, where revisiting successful experiences strengthens memory, consolidates skills, and improves future performance. Motivated by this biological principle, we introduce Instant Episode Repetition (...
42. CORAM: Coherent Orthogonal Rotation for Model Merging ​
Author: Xinyi Sui, Ziran Liu, Nam Ling, Wei Wang, Wei Jiang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17366v1 Announce Type: new Abstract: Merging finetuned models combines specialized capabilities without joint training or access to the original data. Most methods operate by linear arithmetic in Euclidean weight space, which cannot carry the geometry of the update. Orthogonal Model Mergi...
43. Pathology Transport: Optimal-Transport Explanations for Clinical Data, and When Their Heatmaps (Fail to) Localize Disease ​
Author: Lalit Kumar
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17370v1 Announce Type: new Abstract: Generative models promise a route to explainable clinical AI: rather than probe a classifier, model the distributions of healthy and diseased patients and read explanations off the geometry between them. We build such a system - an optimal-transport re...
44. Integrating Novelty and Surprise for Experience Prioritization and Exploration in Image-Based Reinforcement Learning ​
Author: Hoda Yamani, Henry Williams, Bruce A. MacDonald
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17373v1 Announce Type: new Abstract: Sample efficiency is a central challenge in reinforcement learning (RL), particularly in image-based domains where agents must learn from high-dimensional visual inputs. Traditional sampling often relies on random or suboptimal experience selection, le...
45. GUPO: Gradient Uncertainty-aware Policy Optimization for Post-Training Large Language Models ​
Author: Peizheng Guo, Jianqi Zhang, Xingyu Zhang, Yun Fan, Jiahuan Zhou, Changwen Zheng, Wenwen Qiang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17411v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a widely used approach for post-training Large Language Models (LLMs) for reasoning. In GRPO, the group gradients induced by different queries within the same mini-batch are directly averaged to form...
46. General Semantic Knowledge Infusion for Spatio-Temporal Traffic Forecasting ​
Author: Mattis thor Straten, Yannick Wolker, Steffen Strohm, Prathvish Mithare, Ralf Krestel, Matthias Renz
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17440v1 Announce Type: new Abstract: Although Graph Neural Networks (GNNs) have made significant advances in spatio-temporal traffic forecasting, their performance is limited when they rely solely on sensor proximity or road-network topology. This paper presents a spatio-temporal predicti...
47. Causal Local States: Scalable Simultaneous Causal Network Inference and Forecasting for Dynamical Systems ​
Author: Jonas Braun, Fabian Fischbach, Daniel K"oglmayr, Sebastian Baur, Christoph R"ath
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17452v1 Announce Type: new Abstract: Machine learning methods predict many real-world systems with remarkable accuracy, but they are typically treated as black boxes that offer no insight into which interactions drive the dynamics. Causal discovery methods reconstruct the interaction netw...
48. Evaluating RL Explainability Methods by How Much They Help Fix Bugs in Agents ​
Author: Ram Rachum, Yotam Amitai, B'alint Gyevn'ar, Reuth Mirsky, Cameron Allen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17524v1 Announce Type: new Abstract: This preliminary paper outlines a planned evaluation benchmark for Explainable Reinforcement Learning (XRL) methods. Current evaluations rely on functionally-grounded metrics like faithfulness and compactness, and on human-grounded proxies like subject...
49. No Gaussian Required: Contrastive Inverse Dynamics for JEPA World Models ​
Author: Jack Boylan, Chris Hokamp
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17542v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) learn world models by predicting future embeddings, but the objective admits a trivial solution of a constant encoder, so every practical system adds an anti-collapse mechanism (LeCun, 2022; Assran et al...
50. Domain-Adapted Molecular Language Models for Efficient Search of Make-on-Demand Libraries ​
Author: Henrik Wille, Luis-Finley Sch"utz, Felix Strieth-Kalthoff
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.17567v1 Announce Type: new Abstract: Pretrained molecular language models are increasingly used as molecular encoders for learning structure-property relationships. However, their practical suitability for molecular discovery within and beyond their pretraining domain remains unclear. Her...
51. OOD Detection for EEG-based Machine Learning in High-Risk Environments ​
Author: Philipp Bomatter, Henry Gouk
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17620v1 Announce Type: new Abstract: Machine learning models for electroencephalography (EEG) analysis show great promise across a wide range of applications, but their deployment in high-risk domains is hindered by their vulnerability to distribution shifts. Encountering out-of-distribut...
52. rl-triton: High-Performance Triton GPU Kernels for Reinforcement Learning Credit Assignment ​
Author: Lars Simon Zehnder
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.PF
arXiv:2608.17641v2 Announce Type: new Abstract: We present rl-triton, an open-source library of high-performance GPU kernels for reinforcement learning credit assignment, implemented in Triton. The core contribution is a unified associative scan framework that recasts seven distinct RL estimation al...
53. Elimination Geometry ​
Author: Mian Huang, Xueqin Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17646v1 Announce Type: new Abstract: This monograph develops elimination geometry (EG), a typed, native-loss, audit-oriented framework for studying when locally optimal objects can be realized by a shared deployment rule. Elimination and compression may erase distinctions required by pred...
54. Picard Proximal Monte Carlo for Parallel Bayesian Imaging with Score-Based Generative Priors ​
Author: Deliang Wei, Evan Bell, Wenhan Guo, Yifan Chen, Yu Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17666v1 Announce Type: new Abstract: Bayesian imaging inverse problems often require sampling from high-dimensional posterior distributions. While recent score-based and diffusion models provide expressive Bayesian priors, their sampling procedures remain inherently sequential and computa...
55. Conformal Prediction for Molecular Properties under Label Shift ​
Author: Hyeonsu Lee, Juyeon Kim, Erkhembayar Jadamba, Seungjin Choi, Hyunjin Shin
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17678v1 Announce Type: new Abstract: Drug discovery and development underpins healthcare but remains costly and failure-prone. A critical bottleneck lies in predicting molecular properties such as solubility, potency, and toxicity, which directly determine whether a candidate can advance ...
56. Cross-View Correspondence Is a Measurement Intervention: Two-Sided Validation for Agent Evaluation and Credit Assignment ​
Author: Zhen Zhang, Ahmad Hafez, Amr Alanwar
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17713v1 Announce Type: new Abstract: Agent evaluations and trace-based learning often compare outputs across transformed views through a post-response correspondence treated as neutral preprocessing. We show that this correspondence is a measurement intervention: omitting it can manufactu...
57. MAGPIE-Net: Predicting short-duration heavy-rainfall events in station neighborhoods from multitemporal FY-4A AGRI observations ​
Author: Xiang Lin, Yunying Li, Chengzhi Ye, Zitong Chen, Jing Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17753v1 Announce Type: new Abstract: Short-duration heavy-rainfall warning determines whether 1 h rainfall will exceed a threshold within a target-station neighborhood over the next few hours. Multitemporal infrared and water-vapor observations from the Fengyun-4A Advanced Geostationary R...
58. Training-Free Human-in-the-Loop Anomaly Detection via Memory Bank Correction ​
Author: Ayusha Abbas, Saram Abbas, Kabita Adhikari
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17775v1 Announce Type: new Abstract: Anomaly detectors are hardest to deploy exactly where training data is scarcest: a newly commissioned production line has a handful of verified "golden" samples and no machine-learning engineer on the factory floor. We present a training-free human-in-...
59. Debate Training Reduces Reward Hacking in RLAIF ​
Author: Zachary Kenton, Lili Janzer, Rory Greig, Tian Huey Teh, Kirill Tyshchuk, Jonah Brown-Cohen, Harri Edwards, Senthooran Rajamanoharan, Noah Y. Siegel, Natasha Jaques, Rohin Shah
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17776v1 Announce Type: new Abstract: We demonstrate that RL finetuning an LLM using debate, a two-player adversarial game between a generator and a critic adjudicated by a weaker LLM judge, reduces reward hacking compared to a reinforcement learning from AI feedback (RLAIF) baseline. Rewa...
60. Fourth-Moment Geometry of Rademacher Sums ​
Author: Peigan Gao, Jian Qian
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.PR
arXiv:2608.17802v1 Announce Type: new Abstract: Let $\varepsilon_1,\ldots,\varepsilon_n$ be independent Rademacher signs and let $a=(a_1,\ldots,a_n)\in\R^n$ satisfy the normalization below. For the normalized Rademacher sum, we determine how its higher moments depend on the fourth-order mass. Combin...
61. An Empirical Study of Reward Specification and Benchmark Reliability in GRPO-based LLM Unlearning ​
Author: Rub'en Balbastre, Juan Manuel Ordu~na, Mariano P'erez
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.17804v1 Announce Type: new Abstract: Practical LLM unlearning is usually evaluated through two objectives: suppress target-specific knowledge and preserve non-target utility. In generative QA, this leaves a third behavior underspecified: when a target-adjacent prompt admits a broader answ...
62. MotoSafety: Edge-AI with Learned Temporal Importance for Two-Wheeler Collision Risk Assessment Under Time Pressure ​
Author: Sumit S. Shevtekar, Chandresh K. Maurya, Gourab Sil, Subasish Das
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC
arXiv:2608.17823v2 Announce Type: new Abstract: Powered two-wheeler riders face critical safety challenges in low- and middle-income countries, yet limited studies exist on how cognitive stressors such as Time Pressure influence collision risk. We address this gap by introducing a comprehensive data...
63. Leveraging Association Context Retrieval in Knowledge Edit- ing to Build White-Box Attacks on LLMs ​
Author: Roman Maksimov, Vladimir Aletov, Vladimir Solodkin, Dmitry Bylinkin, Daniil Medyakov, Aleksandr Beznosikov
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17836v1 Announce Type: new Abstract: As large language models (LLMs) are granted increasing autonomy, it is essential to investigate methods that can induce unsafe behavior. We propose a novel white-box attack inspired by locate-then-edit approaches from the field of Knowledge Editing. Ou...
64. MoRAX: Mobility-based Representation Augmentation for Geospatial Foundation Models ​
Author: Ya Wen, Jixuan Cai, Yulun Zhou, Alec Kirkley
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2608.17848v1 Announce Type: new Abstract: Geospatial Foundation Models (GFMs) are emerging as a powerful paradigm for learning semantically rich and geographically consistent visual and physical representations. However, their reliance on Earth-observation (EO) data leaves information about hu...
65. Efficient Resource Optimization for Split Federated Learning ​
Author: Wei Wei, Xianhao Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17849v1 Announce Type: new Abstract: Split federated learning (SFL) has emerged as a powerful paradigm for model training at the edge. However, SFL inherently involves discrete decision variables for model splitting and resource allocation, resulting in a challenging mixed-integer problem...
66. Dynamic Compression in Recurrent Networks ​
Author: Jyothish Pari, Ryan Bahlous-Boldi, Pulkit Agrawal
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17896v1 Announce Type: new Abstract: Recurrent models process long contexts efficiently by compressing their history into a fixed-size state, but modern architectures typically do so in a single causal pass over the sequence. Each input must therefore be compressed before the model knows ...
67. Hybrid ML for Lightweight Pre-Route Delay Estimation in Open-Source IC Design ​
Author: Marvin Castro Castro, Erick Carvajal Barboza
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17914v1 Announce Type: new Abstract: Static Timing Analysis (STA) is a critical step in the design flow of digital integrated circuits, however, obtaining accurate delay estimations can represent a challenge when limited information regarding physical design is available. In response, thi...
68. Efficient RLVR Scheduling via Graph-Structured Online Difficulty Estimation ​
Author: Zhizhao Liu, Zhiliang Tian, Xi Wang, Zhihua Wen, Yihang Xiong, Zhiquan Lai, Dongsheng Li
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.17941v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves the reasoning capabilities of large language models but relies on costly rollout exploration. Assigning the same exploration budget to samples with different difficulty levels is inefficien...
69. SIGMA: SHAP-Guided Implicit-Trajectory Generation for Metadata-Free LLM-Based AutoFE ​
Author: Xuan Zheng, Kento Uchida, Shinichi Shirakawa
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.17948v1 Announce Type: new Abstract: Recent research has leveraged Large Language Models (LLMs) to enhance Automated Feature Engineering (AutoFE) through semantic descriptions and trajectory-based prompting. However, there exist two challenges that limit their applicability and scalabilit...
70. An Omitted Mode Is a Rare Rule: The Sampling-Verification Danger Law in Continuous Code World Models ​
Author: Javier Aguilar Mart'in
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY
arXiv:2608.17956v1 Announce Type: new Abstract: In the Code World Model paradigm an LLM synthesizes an executable world model that a classical planner searches, and the model is accepted when it reproduces sampled transitions. We ask what that acceptance certifies in continuous control. We define th...
71. Understanding the Surprising Generalization Properties of Tabular Foundation Models ​
Author: Nour Shaheen, Junwei Ma, Alex Labach, Frank Hutter, Valentin Thomas, Anthony L. Caterini
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17957v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) increasingly rely on in-context learning, where a model receives labelled examples at inference time and predicts labels for new inputs without updating its weights. Existing TFMs are typically trained on either massive...
72. Too Sure to Be Safe: Model Calibration for Reliable Log Anomaly Detection ​
Author: Bin Li, Dongdong Wang, Siyang Lu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE
arXiv:2608.17965v1 Announce Type: new Abstract: Online log anomaly detection is critical for maintaining the reliability of large-scale computing systems. Although recent language model-based log anomaly detectors achieve strong detection performance, their confidence estimates remain poorly calibra...
73. Evaluating and improving crop-yield forecasting methods during extreme drought ​
Author: Shrey Gupta, Yi Ming, George Mohler
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17971v1 Announce Type: new Abstract: The impact of climate variability on food production has led to the creation of various forecasting models that uses machine learning (ML), numerical weather predictors (NWP) or a hybrid of ML-NWP models to identify structural and physical relationship...
74. Recirculation ​
Author: Michael C. Mozer, Shoaib Ahmed Siddiqui, Danny Sawyer, Sunny Sanyal, Rosanne Liu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.17981v1 Announce Type: new Abstract: We describe an inference-time architectural enhancement for off-the-shelf foundation models that markedly reduces perplexity and boosts accuracy across generation and reasoning tasks. Our approach incurs essentially no additional latency during generat...
75. Composing Flow-Matching Energies with Known Physics: Generation, OOD Detection, and Inversion on PDE Fields ​
Author: Yixuan Sun, Anirban Samaddar, Sandeep Madireddy
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2608.18004v1 Announce Type: new Abstract: Probabilistic modeling of physical fields benefits from both a data-driven prior and known physical structure such as the governing equations. Energy-based models (EBMs) are a natural fit since energies compose additively, which enables augmenting phys...
76. Policy-Invariant Reward Shaping from LLM Feedback: A Framework for Hybrid RL Agents ​
Author: Christophe D. Hounwanou, John Emeka Eze, Ya'e U. Gaba
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.18008v1 Announce Type: new Abstract: Combining large language models with reinforcement learning is increasingly explored, yet the theoretical status of LLM-derived reward signals is often left implicit. We formalize the hybrid LLM-planner and RL-controller architecture as a Goal-Augmente...
77. Revisiting WEASEL 2.0: Reproduction, Sensitivity, and an Adaptive Ensemble-Size Rule ​
Author: Cian Higgins, Gerard Carrigan, Pinar Sungu Isiacik, Georgiana Ifrim
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.18021v1 Announce Type: new Abstract: WEASEL 2.0 is a dictionary-based time series classifier that combines dilated sliding windows with a randomised hyperparameter ensemble and a fixed-size dense feature representation. Two of its hyperparameter choices, the maximum ensemble size and the ...
78. Why GPT-Style Models Do Not Directly Transfer to Symbolic Music: Compression in the Wrong Coordinate System ​
Author: Yi Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SD
arXiv:2608.18025v1 Announce Type: new Abstract: GPT-style models achieve strong performance by representing language with finite vocabularies of reusable discrete tokens. This success has motivated symbolic music tokenizations to treat recurring musical structures, such as chords, motifs, and phrase...
79. TabNSM: Neural Sparse Mixer for Tabular Regression ​
Author: Ali Eslamian, Qiang Cheng
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2608.18026v1 Announce Type: new Abstract: Large-scale, high-dimensional tabular regression remains challenging: tree-based models are robust but lack end-to-end representation learning, while deep models enable flexible feature learning but often incur costly interaction modeling and sensitivi...
80. Optimize Your Sampling: Tuned Diffusion Sampling with Bayesian Optimization ​
Author: Travis Zhang, Christian Belardi, Justin Lovelace, Jin Peng Zhou, Saebyeol Shin, Carla P. Gomes, Kilian Q. Weinberger
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.18040v1 Announce Type: new Abstract: Sampling from a diffusion model typically requires many forward passes through a large neural network, making generation computationally expensive. While much work has focused on efficient solvers and samplers, comparatively little attention has been p...
81. The concentration game: Bayesian updating, regret, and information ​
Author: Akshay Balsubramani
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.GT, math.PR, math.ST, stat.TH
arXiv:2608.18061v1 Announce Type: new Abstract: We give a two-player zero-sum repeated game between a learner and nature whose value identity generates Bayesian updating and an exact accounting of exponential-weights regret at once, and supplies the comparator-class variational form that a wide clas...
82. Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs ​
Author: Christos Koutsiaris
Published: 8/20/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG
arXiv:2602.14784v1 Announce Type: cross Abstract: Breaking long documents into smaller segments is a fundamental challenge in information retrieval. Whether for search engines, question-answering systems, or retrieval-augmented generation (RAG), effective segmentation determines how well systems can...
83. Information Spreading in Diffusion Models from Effective Field Theory ​
Author: Navonil Neogi, Nabil Iqbal
Published: 8/20/2026, 4:00:00 AM
Categories: hep-th, cond-mat.stat-mech, cs.LG
arXiv:2608.14308v1 Announce Type: cross Abstract: We study score-matching diffusion models with a convolutional architecture. We argue that the inductive bias of locality means that the machinery of effective field theory from physics can be usefully applied to describe the denoising dynamics. We ap...
84. Advancing Health Equity through Multi-Level Fairness in Health Informatics ​
Author: Nick Souligne, Vignesh Subbian
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CY, cs.LG
arXiv:2608.16902v1 Announce Type: cross Abstract: The increasing integration of machine learning in healthcare has highlighted critical challenges related to fairness, transparency, and health equity. Specifically, the use of multi-level fairness techniques, which combine multiple bias mitigation st...
85. ComNetX: Local Hierarchical Adaptation for Dynamic Community Detection ​
Author: Aleksandr Konovalov, Anna Uporova, Alexander Drobyshev, Iaroslav Egorov, Grigoriy Bokov
Published: 8/20/2026, 4:00:00 AM
Categories: cs.SI, cs.AI, cs.LG
arXiv:2608.16906v1 Announce Type: cross Abstract: Dynamic community detection is commonly addressed either by full-snapshot recomputation or by solver-specific dynamic procedures. Full recomputation preserves the semantics of mature static solvers, but it repeatedly processes unchanged graph regions...
86. Which CS1 Students Will Fail? Identifying Digital Markers from Learning Analytics in Computer Systems and Architecture Using Weighted Academic Momentum and Interaction Logs ​
Author: Lighton Phiri, Mutune Chaibela, Ivy Chisha, David Pungwa, Danny Siabbaba, Bydon Simukoko
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CY, cs.LG
arXiv:2608.16914v1 Announce Type: cross Abstract: Digital learning platforms generate rich behavioural traces (digital markers) that offer the potential to identify struggling students early. This paper investigates whether a combination of traditional and digital markers can predict failure in a fi...
87. MITRE-SAGE: A Multi-Agent Cybersecurity Question-Answering Model ​
Author: Ali Habibzadeh, Farid Feyzi, Reza Ebrahimi Atani
Published: 8/20/2026, 4:00:00 AM
Categories: cs.IR, cs.CR, cs.LG, cs.MA
arXiv:2608.16921v2 Announce Type: cross Abstract: Effective cybersecurity operations require timely and accurate analysis of large-scale heterogeneous security information; however, analysts increasingly struggle with information overload, alert fatigue, and time-constrained decision-making. Althoug...
88. Network Denoising Revisited: A Ricci-Flow-Inspired Graph Diffusion Method ​
Author: Ye Fang, Chuan-Xian Ren
Published: 8/20/2026, 4:00:00 AM
Categories: cs.SI, cs.LG
arXiv:2608.16923v1 Announce Type: cross Abstract: Networks provide a fundamental representation of relationships among entities. However, real-world networks are often corrupted by noise caused by measurement errors and inherent stochasticity, hindering the discovery of meaningful structure. Most de...
89. SPSA Hyperparameter Tuning for Variational Quantum Natural Language Inference ​
Author: Nayan D'Souza, Christopher J. Agostino
Published: 8/20/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.16939v1 Announce Type: cross Abstract: Training variational quantum models requires choosing between parameter-shift gradients, which are exact but cost $O(P)$ forward evaluations, and simultaneous perturbation stochastic approximation (SPSA), which uses only two samples but produces high...
90. A Constant-Competitive Algorithm for Dynamic Mixture-of-Experts Serving ​
Author: Ian D'Ambrosio (Nth Research Collective)
Published: 8/20/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2608.16947v1 Announce Type: cross Abstract: Huang, Lou, and Xiao introduced Dynamic Mixture-of-Experts Serving and gave an O(sqrt(log k))-competitive randomized algorithm for its integral primal problem, where k is the number of replica GPUs beyond the mandatory copy of each expert. Their matc...
91. WONDER: A Radio World Model-based Negotiation Framework for Multi-Agent UAV Coverage Optimization ​
Author: Jiahao Huang, Rongpeng Li, Zhifeng Zhao, Guoru Ding, Honggang Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2608.16955v1 Announce Type: cross Abstract: Post-disaster damage to terrestrial infrastructure can disrupt wireless coverage,while Uncrewed Aerial Vehicle (UAV) swarms provide a promising solution for rapid restoration.However, due to the limitations in local geometry observations hidden radio...
92. The Price of Thinking: Reasoning Effort as a Model-Specific API Contract ​
Author: Yeabin Moon
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CY, cs.LG
arXiv:2608.16956v1 Announce Type: cross Abstract: API buyers purchase a dated contract, not a model name alone: the contract includes the requested and served model, reasoning-effort term or its omission, output rail, service product, prompt, and price schedule. We study the reasoning-effort term th...
93. MagViT: Interpretable Multi-Magnification Transformers with Patient-Level Model Selection for Breast Histopathology ​
Author: Nabil Ashab, Soumit Kumar Kundu, Saif Mahmud Parvez, Shahadat Hossain Sohag, Bidhan Biswas, Nazmus Subha
Published: 8/20/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2608.16959v1 Announce Type: cross Abstract: Breast cancer is one of the most common types of cancer among women around the world. Rapid detection and early treatment can hinder its progress to more complex stages and can impede its spread to other parts of the body. Histopathological image cla...
94. Diagonal Multi-omics Integration of Heterogenous Datasets ​
Author: Maksim V. Kukushkin, Mikhail S. Arbatskiy, Dmitriy E. Balandin, Alexey V. Churov
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.FA
arXiv:2608.16968v1 Announce Type: cross Abstract: In this paper, we consider methods for the diagonal multi-omics integration of heterogeneous datasets. Several approaches to the nature of biological heterogeneity are analyzed and developed to comprehend more clearly the generated differences. Speci...
95. Probing the Prefill: Detecting Code Vulnerabilities via Latent Activations ​
Author: Alizishaan Khatri
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2608.16970v1 Announce Type: cross Abstract: LLM-based code generation is now embedded in mission-critical pipelines, but defenses against vulnerable output remain post-hoc -- static analyzers, fine-tuned classifiers, or an LLM judge that screen completed code, ignoring the generating model's o...
96. FedPref: Federated Preference Learning for Structured Radiology Report Extraction ​
Author: Flint Xiaofeng Fan, Cheston Tan, Yew-Soon Ong, Roger Wattenhofer
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.16971v1 Announce Type: cross Abstract: Radiology reports describe findings and locations in free text, but downstream search and analysis require these relations in a fixed schema. Learning this extraction requires labels that are unevenly distributed across institutions: smaller hospital...
97. Margin-Regularized Structured Semantic Alignment for Brain-Language Correspondence ​
Author: Jiaqi Wang, Huawen Hu, Shu Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.16975v1 Announce Type: cross Abstract: With the rapid advancement of large language models, brain-language decoding has achieved remarkable progress. However, it remains unclear whether decoded content genuinely reflects neural representations or is largely reconstructed by the language m...
98. VLCP: Vision Language Control Policy Closed-Loop Code Replanning for Robot Manipulation ​
Author: Dhia Naouali, Minghan Wu, Claudia Wong, Abhinav Puthran, Omar G. Younis
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.16978v1 Announce Type: cross Abstract: Turning a frontier vision-language model into a robot policy usually means fine-tuning it to emit an action representation it never saw in pretraining, which throws away much of the reasoning that made the model worth reaching for. We go the other wa...
99. Lambda-Hold Control: Human-Like Movement Emerges from a Minimal Task Reward in Predictive Musculoskeletal Simulation ​
Author: Jun Hyuk Lee, Chihyeong Lee, Jooeun Ahn
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.GR, cs.LG
arXiv:2608.17030v1 Announce Type: cross Abstract: The massive overactuation in the human musculoskeletal system makes it challenging to train musculoskeletal models to generate human-like motion via reinforcement learning, primarily because exploration in the resulting high-dimensional and redundant...
100. Wasted large language models: A life cycle thinking approach ​
Author: Erik Johannes Husom, Maria Emine Nylund, Ophelia Prillard
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CY, cs.HC, cs.LG
arXiv:2608.17055v1 Announce Type: cross Abstract: Large Language Models (LLMs) are machine learning (ML) models that have an increasingly large carbon footprint through their development and use. Efforts to increase the energy efficiency of these models have not translated into reduced consumption d...
101. Dynamic Entanglement-Weighted Pruning for Quantum Federated Unlearning in Supply-Chain Risk Prediction ​
Author: Aditya Kumar, Sumit Chongder
Published: 8/20/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.17069v1 Announce Type: cross Abstract: Federated deployments of variational quantum classifiers are attractive for cross-organisation risk prediction in supply chains, because raw data never leaves the client, yet data-protection regulations such as the GDPR grant clients a right to reque...
102. Digital Twin-Based Intrusion Detection for Vehicle Powertrain CAN Bus Systems ​
Author: Araf Rahman, M Sabbir Salek, Mashrur Chowdhury
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.17093v1 Announce Type: cross Abstract: Existing automotive intrusion detection systems (IDSs) for the Controller Area Network (CAN) largely target discrepancies in message timing, frequency, or sequencing and cannot detect attacks that preserve these properties while manipulating the payl...
103. Picture the Epsilon: Pursuing Identity-Level Privacy Guarantees for Images ​
Author: Arman Zareian Jahromi, Vishnu Bondalakunta, Mohammad Akbar Bin Shah, Naimul Haque, Shuangqing Wei, George T. Amariucai
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.17147v1 Announce Type: cross Abstract: Image-to-image face generators are widely used, and visual dissimilarity between their outputs and source images is sometimes treated as evidence of privacy. Auditing whether these systems satisfy formal identity-level (epsilon, delta)-differential p...
104. Lymphocyte Mimicry Correction via Region-Level Tissue Reasoning and Unbalanced Optimal Transport ​
Author: Xiang Li, Yuqi Wang, Casey C. Heirman, Jihye Heo, Kyle J. Lafata
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.17151v1 Announce Type: cross Abstract: Cell mimicry arises when different cell types appear morphologically similar. Human pathologists resolve this ambiguity using surrounding tissue context, whereas current vision models either lack contextual reasoning (cell foundation models) or canno...
105. Policy Optimization and Statistical Inference for Online Contextual Matrix Games ​
Author: Liner Xiang, Yixin Wang, Hengrui Cai
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.ME, stat.TH
arXiv:2608.17173v1 Announce Type: cross Abstract: Online decision making often requires navigating a landscape shaped by both dynamic contexts and strategic interactions. In competitive pricing, for example, hotels must account for both dynamic contextual factors and rivals' strategic responses. Exi...
106. Expressivity In Multimodal Contrastive Learning ​
Author: Andrew Stuart, Florian Wolf
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2608.17203v1 Announce Type: cross Abstract: Contrastive learning has become a cornerstone of modern representation learning, powering CLIP-style models that underpin text-to-image generation, vision-language models, and retrieval across a rapidly growing range of modalities. Despite this empir...
107. Teach and Grow: An Agent-Centered Architecture for General Robot Learning ​
Author: Chang Nie, Zhe Liu, Hesheng Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG
arXiv:2608.17209v1 Announce Type: cross Abstract: End-to-end vision-language-action (VLA) and world-action models offer an elegant route to general-purpose robotics, but their reliability is bounded by validated physical coverage. When an unfamiliar object, sensor, embodiment, or contact falls outsi...
108. Temporal Leakage in Financial News NLP: A Multi-Architecture Audit with a Regime-Specific M&A Signal ​
Author: Chenhao Xue, Raslen Guesmi, Siwei Feng, Yucheng Gong, Jacob Xavier Sundram, Jordan Pang, Lan Wang, Julian Kaljuvee
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.17223v1 Announce Type: cross Abstract: Financial-news direction prediction has become a popular NLP benchmark, yet reported gains depend critically on whether the train-test split is chronological or random, i.e., on temporal leakage. We audit this dependence on a 49,799-article corpus ac...
109. Information fusion and machine learning for sensitivity analysis using physics knowledge and experimental data ​
Author: Berkcan Kapusuzoglu, Sankaran Mahadevan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CE, cs.LG, stat.ME, stat.ML
arXiv:2608.17248v1 Announce Type: cross Abstract: When computational models (either physics-based or data-driven) are used for the sensitivity analysis of engineering systems, the sensitivity estimate is affected by the accuracy and uncertainty of the model. This paper considers global sensitivity a...
110. Adaptive surrogate modeling for high-dimensional spatio-temporal output ​
Author: Berkcan Kapusuzoglu, Shunsaku Matsumoto, Yoshitomo Miyagi, Daigo Watanabe, Sankaran Mahadevan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CE, cs.AI, cs.LG, stat.ME, stat.ML
arXiv:2608.17250v1 Announce Type: cross Abstract: This paper develops an adaptive surrogate modeling method for problems with very high-dimensional spatio-temporal outputs. The analysis of spatio-temporal multi-physics systems is computationally expensive and consists of a large number of inputs and...
111. SPACE: Sample-cloud Predictive Adaptive Conformal Ellipsoids for Multivariate Time-Series Forecasting ​
Author: Baishi Li, Kelvin J. L. Koa, Ke-Wei Huang
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2608.17333v1 Announce Type: cross Abstract: Modern probabilistic time-series forecasters often express uncertainty through forecast samples. While typically converted into nominal prediction regions using empirical quantiles, these model-implied sets lack formal coverage guarantees and frequen...
112. On the Pseudo-Mixing of Kac's Walk ​
Author: Natesh S. Pillai, Aaron Smith, Vinod Vaikuntanathan
Published: 8/20/2026, 4:00:00 AM
Categories: math.PR, cs.CR, cs.LG
arXiv:2608.17374v1 Announce Type: cross Abstract: Motivated by a conjecture of Vaikuntanathan and Zamir, we study the pseudo-mixing of Kac's walk on $\mathrm{SO}(n)$: whether short trajectories are indistinguishable from Haar measure by low-complexity tests. We prove that the first $k$ columns mix i...
113. Leveraging generative hallucination and biophysics-informed modeling for unified biomolecular sequence-structure co-design ​
Author: Xuefeng Liu, Mingxuan Cao, Xiao Luo, Songhao Jiang, Tobin Sosnick, Jinbo Xu, Louis Maher, Rick Stevens
Published: 8/20/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG, q-bio.BM
arXiv:2608.17381v1 Announce Type: cross Abstract: Biomolecular design underpins applications from molecular recognition to therapeutics and synthetic biology, yet de novo interaction design remains challenging-especially for DNA/RNA, underexplored non-protein modalities with scarce, heterogeneous co...
114. Prism-GRPO: Faster VLA Policy Optimization via Splitting Same-outcome Groups ​
Author: Zeyun Deng, Yuzhe Lu, Yawei Wang, Linbo Liu, Qing Ping, Han Ding, Guande Wu, Panpan Xu, Jun Huan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.17423v1 Announce Type: cross Abstract: GRPO is increasingly used for reinforcement learning of vision-language-action (VLA) policies because, unlike PPO, it does not require training a critic. This simplification comes with a sampling cost: group-relative advantages require multiple rollo...
115. Nonlocal Transition Kernel for Efficient Learning of Restricted Boltzmann Machines ​
Author: Kaiji Sekimoto, Muneki Yasuda
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.dis-nn, cs.LG
arXiv:2608.17450v1 Announce Type: cross Abstract: Learning restricted Boltzmann machines (RBMs) is computationally challenging because it requires expectations whose exact evaluation is generally intractable. The expectations are typically evaluated using a sampling approximation based on blocked Gi...
116. Online Generalized Sparse Regression: How Does Overparametrization Help? ​
Author: Shuoguang Yang, Qiang Sun
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2608.17466v1 Announce Type: cross Abstract: Regularized sparse regression has been extensively studied in the offline setting, but online formulation remains relatively under-explored. This gap stems from four key challenges: (i) the infeasibility of dynamically updating the regularization par...
117. When AI Designs AI: Innovation or Imitation? ​
Author: Yikang Yang, Zhengxin Yang, Luzhou Peng, Minghao Luo, Yanqi Kan, Wanling Gao, Jianfeng Zhan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.17471v1 Announce Type: cross Abstract: Recent advances in LLM agents have made them increasingly capable of designing methods for complex AI tasks. This raises two central questions about agent-designed methods relative to human-designed methods: how well they perform, and how different t...
118. SGHA: Evidence-Grounded Research Problem Discovery with Local Language Models ​
Author: Sarvesh Gharat, Junpei Komiyama
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.17501v1 Announce Type: cross Abstract: Recent efforts toward fully automated AI scientists have demonstrated that language-model agents can generate hypotheses, execute experiments, and draft scientific manuscripts. However, during the early stages of research, when research problems are ...
119. Looking Beyond the Scale: Do Surgical Skill Models Learn Transferable Representations Across Assessment Rubrics? ​
Author: Hanna Hoffmann, Felix von Bechtolsheim, Stefanie Speidel, Rebecca Hisey
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.17519v1 Announce Type: cross Abstract: Vision-based surgical skill assessment has shown strong in-domain results, yet a fundamental question remains unasked: do these models learn transferable representations of surgical proficiency, or do they merely encode dataset-specific visual patter...
120. When to Review: Spaced Repetition for Continual Pre-Training of Language Models ​
Author: Alankar Atreya, Devesh Batra, Yoages Kumar Mantri, Geremy Bantug, Greig A Cowan, Raad Khraishi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.17530v1 Announce Type: cross Abstract: Continual pre-training of large language models must acquire new information without erasing old knowledge. Existing replay methods often choose a global old/new mixture and sample uniformly, ignoring that examples differ in how quickly they are forg...
121. Reflex-Guard: A Low-Latency Guardrail for LLM Prompt Safety Using Dense Semantic Embeddings ​
Author: Istiaque Ahmed, Afia Anjum Borsha, Ranat Das Prangon, Abu-fuad Ahmad, Thi Hong Tran
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG
arXiv:2608.17556v1 Announce Type: cross Abstract: Large Language Models (LLMs) in real-world applications often face the risks of specially crafted prompts designed to bypass the safety controls. Existing guardrail methods, such as LLM-as-a-judge and cloud-based safety APIs are able to detect unsafe...
122. Leveraging existing sparse point annotations for benthic imagery dense segmentation ​
Author: Cesar Borja, Breck A. McCollum, Jarret E. Byrnes, Kenneth Sebens, Ana C. Murillo
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.17561v1 Announce Type: cross Abstract: The health of marine ecosystems is a critical indicator of global environmental change, yet the physical constraints of underwater observation and the intrinsic challenges of processing marine imagery severely limit the scalability of systematic moni...
123. Feature Priming in Online Linear Regression: Sparse-Regret Lower Bounds and a Tight Univariate Rate ​
Author: Huibo Xu, Shi Fu, Qixin Zhang, Dacheng Tao
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP
arXiv:2608.17573v1 Announce Type: cross Abstract: In high-dimensional online prediction, the best predictor may depend on only a few features, so regret should scale with sparsity rather than the ambient dimension. Feature priming pursues this goal by estimating feature weights from past data and re...
124. Communication Reduction via Semantic-Based Encoding in DMPC Using LSTMs ​
Author: Torben Schiz, Pedro H. J. Nardelli, Henrik Ebel
Published: 8/20/2026, 4:00:00 AM
Categories: eess.SY, cs.DC, cs.LG, cs.MA, cs.RO, cs.SY
arXiv:2608.17592v1 Announce Type: cross Abstract: The communication demands of distributed model prediction control (DMPC) can overwhelm even advanced wireless communication technologies as agents must exchange a significant amount of information at least once per time step. To semantically reduce c...
125. MoNe: Modular Neural Memory for Efficient Long Context Inference ​
Author: Wonguk Cho, Kyubyung Chae, Tribhuvanesh Orekondy, Sunghyun Park, Hyoungwoo Park, Jeongho Kim, Arash Behboodi, Kyuwoong Hwang, Sungrack Yun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.17616v1 Announce Type: cross Abstract: We present MoNe, a lightweight modular neural memory that attaches to any frozen pretrained Transformer to enable long-context inference without retraining. MoNe reads context in fixed-size segments via test-time learning of fast-weight neural memory...
126. Iterative Grasp Pose Refinement: A Deep Reinforcement Learning Approach for 2D Vision ​
Author: Amir Arsalan Nematollahi, Shayan Ahmadi, Mehdi Tale Masouleh, Ahmad Kalhor
Published: 8/20/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, eess.IV
arXiv:2608.17628v1 Announce Type: cross Abstract: Developing robots capable of understanding and manipulating objects requires compact, interpretable, and generalizable representations. This work proposes a reinforcement learning-based framework for robotic grasp refinement, integrating keypoint-bas...
127. Mixture-of-Expert Blocks Contain Strong Hallucination Detection Signals ​
Author: Joao Fonseca, Rodrigo Rodrigues, Paolo Romano
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.17687v1 Announce Type: cross Abstract: Despite their widespread use, Large Language Models (LLMs) remain limited by a fundamental problem: the generation of plausible but false content, known as hallucinations. Most existing detection methods operate at the answer or sentence level, yet p...
128. MemCatalyst: Amplifying Data Auditing on Vision-Language Models via Data Poisoning ​
Author: Xukun Luan, Jinyan Liu, Yuhui Gong, Yuanguo Bi, Bing Hu, Xuesong Li, Di Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.17722v1 Announce Type: cross Abstract: Vision-Language models (VLMs) achieve outstanding performance largely due to the amount of training data available on the internet. At the same time, data holders (e.g., artists) urgently need to determine whether their data has been used for model t...
129. Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See ​
Author: Ayoub Kirouane, Christos Petrocheilos
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.RO, stat.ML
arXiv:2608.17744v1 Announce Type: cross Abstract: Take three frontier mixture-of-experts models (Alibaba, OpenAI, NVIDIA; 3.6-4.0B active parameters each) and fine-tune them to reason in a low-resource language. On accuracy benchmarks almost nothing happens, and the benchmark itself is noise at this...
130. Diff-DDoS: Realistic Cyber-Physical Attack Synthesis and Robust Detection for 5G-Enabled CPS Using Tabular Diffusion Models ​
Author: Bilal Hussain, Xiao Tang, Qinghe Du, Tan Li, Muhammad Azhar, Danista Khan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.17796v1 Announce Type: cross Abstract: Deep learning-based DDoS detectors for 5G-enabled cyber-physical systems face scarce labeled attack data and unrealistic synthetic substitutes, which limit robustness against adaptive adversaries. Detectors trained on hand-crafted attacks with fixed ...
131. Spatially explicit feature importance for building height estimation using research-access high-resolution SAR and optical sensors ​
Author: Guilherme Iablonovski, Pierre-Louis Frison, Tatiana Silva da Silva
Published: 8/20/2026, 4:00:00 AM
Categories: physics.soc-ph, cs.LG, stat.ML
arXiv:2608.17822v1 Announce Type: cross Abstract: Accurate building height information at the individual footprint scale is essential for material stock accounting and post-disaster damage assessments yet remains difficult to obtain at city scale in the Global South where airborne LiDAR coverage is ...
132. Toward the Optimal Regret-Instability Trade-off in Multi-Armed Bandits ​
Author: Kaifei Wang, Yinyu Ye, Han Zhong
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC, math.ST, stat.TH
arXiv:2608.17841v1 Announce Type: cross Abstract: Multi-armed bandit algorithms are evaluated by regret, yet comparable regret can coexist with different allocations across independent runs. We study the trade-off between worst-case regret $\mathcal{R}{K,T}$ and instability $\mathcal S$, defi...
133. A Residual Learning Approach for Unsteady Aerodynamic Load Prediction ​
Author: Divya Sanghi, Carlos E. S. Cesnik
Published: 8/20/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG
arXiv:2608.17894v1 Announce Type: cross Abstract: This paper investigates the feasibility of using residual learning to improve unsteady aerodynamic load prediction for aeroelastic applications. The machine learning technique selected for the study is the long short-term memory (LSTM) neural network...
134. AppendiGrade: An XAI-Enhanced Deep Learning Framework for Grading Appendicitis in Ultrasound with Gaussian Blur and Grad-CAM ​
Author: Fahad Ahammed, Omar Faruq Shikdar, Navid Zaman, Md Tahsin, Md. Nawab Yousuf Ali, Golam Sorwar
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.17923v1 Announce Type: cross Abstract: Appendicitis is one of the most common abdominal emergencies worldwide and requires prompt diagnosis and treatment to prevent life-threatening conditions. However, accurately differentiating complicated cases, such as perforation or abscess formation...
135. Procedural Content Metageneration via Program Search and Continual Abstraction Discovery ​
Author: Matthew Siper, Ahmed Khalifa, Julian Togelius
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NE
arXiv:2608.17947v1 Announce Type: cross Abstract: Large language models can generate executable programs, which makes it possible to search directly over procedural content generators rather than individual levels. We study this approach in Sokoban, Zelda, Dangerous Dave, and Lode Runner. Each run e...
136. Towards Zero-Shot Task Transfer with Neurosymbolic World Models ​
Author: Isidoro Tamassia, Lennert De Smet, Giuseppe Marra
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.17959v1 Announce Type: cross Abstract: State-of-the-art model-based reinforcement learning methods learn neural world models that allow policy improvement by planning in a latent space, without assumptions on the structure of the underlying environment. While expressive, these models are ...
137. Against Political Polarization: A Unified Framework for Tracing Evolving Political Ideologies on Social Media ​
Author: Yijie Xu, Chao Wang, Hui Xiong
Published: 8/20/2026, 4:00:00 AM
Categories: cs.SI, cs.AI, cs.CL, cs.LG
arXiv:2608.17987v1 Announce Type: cross Abstract: The rapid growth of social media has greatly influenced political discourse, highlighting the need to understand individual political ideologies and their temporal dynamics. This task faces challenges such as data scarcity, abundant non-political con...
138. Where A Small Language Model Helps in Invoice Categorisation, Understood Through Embedding Geometry ​
Author: Emma Ceccherini, Daniel Lawson, Anjulika Salhan
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.18033v1 Announce Type: cross Abstract: Categorising invoices into the correct General Ledger (GL) code underpins financial reporting and tax compliance. This is a skilled accounting judgement rather than a routine task: the correct category depends subtly on the nature of the purchasing b...
139. Harnessing Magnitude-Only and Complex Measurements for Improved Dynamic MRI Reconstruction with Learned Priors ​
Author: Mahdi Saberi, Ya\c{s}ar Utku Al\c{c}alar, Merve G"{u}lle, Chetan Shenoy, Mehmet Ak\c{c}akaya
Published: 8/20/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG, physics.med-ph
arXiv:2608.18036v1 Announce Type: cross Abstract: MRI reconstruction methods for undersampled k-space data naturally utilize complex-valued measurements. Parallel developments in sparse phase retrieval have shown that magnitude-only measurements may provide complementary information for signal recov...
140. Primitive Representation Learning for Unsupervised Dynamic Contrast Enhanced MRI Reconstruction ​
Author: Veronika Spieker, Wenqi Huang, Cemre Ariyurek, Liam Timms, Daniel Rueckert, Onur Afacan, Julia A. Schnabel, Sila Kurugol
Published: 8/20/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, eess.SP, physics.med-ph
arXiv:2608.18055v2 Announce Type: cross Abstract: Reliable quantitative analysis of dynamic contrast-enhanced MRI requires high-quality spatiotemporal reconstructions at high undersampling rates. Scan-specific reconstructions using Gaussian and Gabor primitives have shown promising results without t...
141. TokEval: A Tokenizer Evaluation Suite ​
Author: Clara Meister
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.18062v1 Announce Type: cross Abstract: Language model tokenizers are typically selected with minimal evaluation, despite the fact that their design choices directly impact model capabilities. This can be partly attributed to a limited understanding of which tokenizer properties affect whi...
142. On the Fragility of Self-Improving Agents: Variance, Task Order, and Underspecification ​
Author: Qinyuan Ye, Yu Li, Yada Pruksachatkun, Jiaxin Zhang, Chien-Sheng Wu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.18066v1 Announce Type: cross Abstract: Memory-based self-improving agents--those that learn from an online stream of tasks and improve over time by maintaining a textual memory bank--have shown great promise in recent literature. However, the reliability aspects of these methods have been...
143. HyPE-GT: where Graph Transformers meet Hyperbolic Positional Encodings ​
Author: Kushal Bose, Swagatam Das
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2312.06576v2 Announce Type: replace Abstract: Graph Transformers (GTs) facilitate the comprehension of complex relationships on graph-structured data by leveraging self-attention of the possible pairs of nodes. The structural information or inductive bias of the input graph is provided as posi...
144. Diffusion Models for Smarter UAVs: Decision-Making and Modeling ​
Author: Yousef Emami, Hao Zhou, Luis Almeida, Kai Li
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2501.05819v2 Announce Type: replace Abstract: Uncrewed Aerial Vehicles (UAVs) are increasingly used in modern communication networks. However, challenges in decision-making and digital modeling continue to hinder their rapid development. Reinforcement Learning (RL) algorithms face limitations ...
145. Gradient Heterogeneity Complements Hessian Heterogeneity in Transformer Optimization ​
Author: Akiyoshi Tomihari, Issei Sato
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2502.00213v5 Announce Type: replace Abstract: Transformers are difficult to optimize with stochastic gradient descent (SGD) and largely rely on adaptive optimizers such as Adam. Despite extensive efforts, the mechanisms behind Adam's advantage over SGD in Transformer optimization are still not...
146. LZ Penalty: An information-theoretic repetition penalty for autoregressive language models ​
Author: Antonio A. Ginart, Naveen Kodali, Jason Lee, Caiming Xiong, Silvio Savarese, John R. Emmons
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT
arXiv:2504.20131v4 Announce Type: replace Abstract: We introduce the LZ penalty, a penalty specialized for reducing degenerate repetitions in autoregressive language models without loss of capability. The penalty is based on the codelengths in the LZ77 universal lossless compression algorithm. Throu...
147. TabularQGAN: A quantum generative model for tabular data synthesis ​
Author: Pallavi Bhardwaj, Caitlin Jones, Lasse Dierich, Aleksandar Vu\v{c}kovi'c
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, quant-ph
arXiv:2505.22533v2 Announce Type: replace Abstract: In this paper, we introduce a novel quantum generative model for synthesizing tabular data. Synthetic data is valuable in scenarios where real-world data is scarce or private, as it can be used to augment or replace existing datasets. As enterprise...
148. Global Convergence of Gradient EM for Over-Parameterized Gaussian Mixtures ​
Author: Mo Zhou, Weihang Xu, Maryam Fazel, Simon S. Du
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2506.06584v2 Announce Type: replace Abstract: Learning Gaussian Mixture Models (GMMs) is a fundamental problem in statistics and machine learning, with the Expectation-Maximization (EM) algorithm and its popular variant gradient EM being arguably the most widely used algorithms in practice. In...
149. Monotone Classification with Relative Approximations ​
Author: Yufei Tao
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2506.10775v3 Announce Type: replace Abstract: In monotone classification, the input is a multi-set $P$ of points in $\mathbb{R}^d$, each associated with a hidden label from ${-1, 1}$. The goal is to identify a monotone function $h$, which acts as a classifier, mapping from $\mathbb{R}^d$ to ...
150. Continuous Evolution Pool: Taming Recurring Concept Drift in Online Time Series Forecasting ​
Author: Tianxiang Zhan, Ming Jin, Yuanpeng He, Yuxuan Liang, Shirui Pan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2506.14790v3 Announce Type: replace Abstract: Recurring concept drift is pervasive in real-world online time series, where the underlying data-generating process repeatedly alternates between a small set of regimes, most notably daily or seasonal cycles that dominate energy, traffic, and weath...
151. Scientific Machine Learning of Chaotic Systems Learns Reduced-Order Equations for Neural Populations ​
Author: Anthony G. Chesebro, David Hofmann, Vaibhav Dixit, Earl K. Miller, Richard H. Granger, Alan Edelman, Christopher V. Rackauckas, Lilianne R. Mujica-Parodi, Helmut H. Strey
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, nlin.CD, q-bio.NC
arXiv:2507.03631v5 Announce Type: replace Abstract: Extracting interpretable mathematical models from complex dynamical systems is difficult, especially for chaotic dynamics observed with noisy experimental data. We present PEM-UDE, a method that combines prediction-error methodology with universal ...
152. Exact Reformulation and Optimization for Direct Metric Optimization in Binary Imbalanced Classification ​
Author: Le Peng, Yash Travadi, Chuan He, Ying Cui, Ju Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2507.15240v2 Announce Type: replace Abstract: For classification with imbalanced class frequencies, i.e., imbalanced classification (IC), standard accuracy is known to be misleading as a performance measure. While most existing methods for IC resort to optimizing balanced accuracy (i.e., the a...
153. Neural Operator-Based Nonlinear Nudging for Chaotic Dynamical Systems ​
Author: Jaemin Oh, Jinsil Lee, Youngjoon Hong
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2508.05778v2 Announce Type: replace Abstract: Nudging is an empirical data assimilation technique that incorporates an observation-driven control term into the model dynamics. The trajectory of the nudged system approaches the true system trajectory over time, even when the initial conditions ...
154. HeteRo-Select: Informativeness as the Participation Driver in Heterogeneous Federated Learning ​
Author: Md. Akmol Masud, Md Abrar Jahin, Mahmud Hasan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2508.06692v3 Announce Type: replace Abstract: Federated learning systems typically allocate gradient compression by link speed. This is sensible when bandwidth and data informativeness align. However, under non-IID data, these signals often decorrelate or invert. A bandwidth-driven allocator t...
155. Estimating Parameter Fields in Multi-Physics PDEs from Scarce Measurements ​
Author: Xuyang Li, Mahdi Masmoudi, Rami Gharbi, Nizar Lajnef, Vishnu Naresh Boddeti
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2509.00203v3 Announce Type: replace Abstract: Parameterized partial differential equations (PDEs) underpin the mathematical modeling of complex systems in diverse domains, including engineering, healthcare, and physics. A central challenge in using PDEs for real-world applications is to accura...
156. Asynchronous Message Passing for Addressing Oversquashing in Graph Neural Networks ​
Author: Kushal Bose, Swagatam Das
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.06777v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) suffer from oversquashing, where structural bottlenecks limit message propagation between distant nodes, hindering tasks that require long-range interactions. Existing remedies are limited: graph rewiring alters edge co...
157. One Pipeline, Many Transformers: Pattern-Specific Imputation Specialists for Tabular Missing Data ​
Author: Jacob Feitelberg, Dwaipayan Saha, Kyuseong Choi, Zaid Ahmad, Anish Agarwal, Raaz Dwivedi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.02625v5 Announce Type: replace Abstract: Missing data in tabular datasets forces practitioners into a hard choice: deploy a general-purpose imputer that may perform poorly for the problem at hand, or wait for someone to design a specialized algorithm. This problem is worsened by the fact ...
158. A Weak Penalty Neural ODE for Learning Chaotic Dynamics from Noisy Time Series ​
Author: Xuyang Li, John Harlim, Dibyajyoti Chakraborty, Romit Maulik
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.DS
arXiv:2511.06609v5 Announce Type: replace Abstract: The accurate forecasting of complex, high-dimensional dynamical systems from observational data is a fundamental task across numerous scientific and engineering disciplines. A significant challenge arises from noisy observations of deterministic dy...
159. Row-Stochastic Matrices Can Provably Outperform Doubly Stochastic Matrices in Decentralized Learning ​
Author: Bing Liu, Boao Kong, Limin Lu, Kun Yuan, Chengcheng Zhao
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.19513v4 Announce Type: replace Abstract: Decentralized learning often involves a weighted global loss with heterogeneous node weights $\lambda$. We revisit two natural strategies for incorporating these weights: (i) embedding them into the local losses to retain a uniform weight (and thus...
160. Cluster Aggregated GAN (CAG): A Cluster-Based Hybrid Model for Appliance Pattern Generation ​
Author: Zikun Guo, Adeyinka. P. Adedigba, Rammohan Mallipeddi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2512.22287v4 Announce Type: replace Abstract: Synthetic appliance data are essential for developing non-intrusive load monitoring algorithms and enabling privacy preserving energy research, yet the scarcity of labeled datasets remains a significant barrier. Recent GAN-based methods have demons...
161. Parametric and Generative Forecasts of EPEX Day-Ahead Energy Market Curves ​
Author: Julian Gutierrez, Redouane Silvente
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2601.20226v3 Announce Type: replace Abstract: We propose two methodologies for modelling aggregated supply and demand curves in the EPEX SPOT Day-Ahead market, emphasizing generative models as a way to recover distributional variability. The first is a low-dimensional parametric representation...
162. How (Not) to Hybridize Neural and Mechanistic Models for Epidemiological Forecasting ​
Author: Yiqi Su, Ray Lee, Jiaming Cui, Naren Ramakrishnan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.06323v3 Announce Type: replace Abstract: Epidemiological forecasting from surveillance data is a hard problem and hybridizing mechanistic compartmental models with neural models is a natural direction. The mechanistic structure helps keep trajectories epidemiologically plausible, while ne...
163. Community Concealment from Graph Neural Networks ​
Author: Dalyapraz Manatova, Pablo Moriano, L. Jean Camp
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.SI
arXiv:2602.12250v2 Announce Type: replace Abstract: Graph neural networks (GNNs) enable powerful unsupervised learning of communities. However, such inference may inadvertently expose sensitive group structures, critical clustered patterns, or collective behaviors, raising concerns about sensitive g...
164. TiMi: Empower Time Series Transformers with Multimodal Mixture of Experts ​
Author: Jiafeng Lin, Yuxuan Wang, Huakun Luo, Jianmin Wang, Zhongyi Pei
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.21693v2 Announce Type: replace Abstract: Multimodal time series forecasting has garnered significant attention for its potential to provide more accurate predictions than traditional single-modality models by leveraging rich information inherent in other modalities. However, due to fundam...
165. Quantifying Memorization and Privacy Risks in Genomic Language Models ​
Author: Alexander Nemecek, Wenbiao Li, Xiaoqian Jiang, Jaideep Vaidya, Erman Ayday
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, q-bio.GN
arXiv:2603.08913v2 Announce Type: replace Abstract: Genomic language models (GLMs) have emerged as powerful tools for learning representations of DNA sequences, enabling advances in variant prediction, regulatory element identification, and cross-task transfer learning. However, as these models are ...
166. How to make the most of your masked language model for protein engineering ​
Author: Calvin McCarter, Nick Bhattacharya, Sebastian W. Ober, Hunter Elliott
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2603.10302v3 Announce Type: replace Abstract: A plethora of protein language models have been released in recent years. Yet comparatively little work has addressed how to best sample from them to optimize desired biological properties. We fill this gap by proposing a flexible, effective sampli...
167. Likelihood Hacking in Probabilistic Program Synthesis ​
Author: Jacek Karwowski, Younesse Kaddar, Zihuiwen Ye, Esmeralda S. Whitammer, Sam Staton
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.PL
arXiv:2603.24126v2 Announce Type: replace Abstract: When language models are trained by reinforcement learning (RL) to write probabilistic programs, they can artificially inflate their marginal-likelihood reward by producing programs whose data distribution fails to normalise instead of fitting the ...
168. Self-Distillation as a Performance Recovery Mechanism for LLMs: Counteracting Compression and Catastrophic Forgetting ​
Author: Chi Liu, Xin Chen, Xu Zhou, Fangbo Tu, Srinivasan Manoharan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2604.15794v2 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved remarkable success, underpinning diverse AI applications. However, they often suffer from performance degradation due to factors such as catastrophic forgetting during Supervised Fine-Tuning (SFT), quantiz...
169. Protect the Brain When Treating the Heart: Feasibility of 2.5D U-Net for Real-Time Gaseous Microemboli Detection ​
Author: Andrea Angino, Ken Trotti, Diego Ulisse Pizzagalli, Rolf Krause, Tiziano Torre, Stefanos Demertzis
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2604.22258v2 Announce Type: replace Abstract: Gaseous microemboli (GME) represent a common complication of cardiac structural interventions across both surgical and transcatheter approaches. Intraoperative transesophageal echocardiography (TEE) represents a convenient methodology to monitor an...
170. The Optimal Sample Complexity of Multiclass and List Learning ​
Author: Chirag Pabbaraju
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2604.24749v4 Announce Type: replace Abstract: While the optimal sample complexity of binary classification in terms of the VC dimension is well-established, determining the optimal sample complexity of multiclass classification has remained open. The appropriate complexity parameter for multic...
171. A Finite-Iteration Theory for Asynchronous Categorical Distributional Temporal-Difference Learning ​
Author: Ege C. Kaya, Abolfazl Hashemi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2605.06866v2 Announce Type: replace Abstract: We study finite-iteration behavior of the exact asynchronous recursions used by categorical distributional temporal-difference methods. The analysis covers scalar categorical TD in the Cram'er geometry and multivariate signed-categorical TD in the...
172. Latent Order Bandits ​
Author: Emil Carlsson, Newton Mwai, Fredrik D. Johansson
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.07304v2 Announce Type: replace Abstract: Bandit algorithms solve diverse sequential decision-making problems, but are often too sample-inefficient for from-scratch personalization. To substantially reduce exploration times, latent bandit algorithms exploit cross-instance structure implied...
173. Center-Manifold Reduction of Learning at Bifurcations: Interference and Rich Learning in Recurrent Neural Networks ​
Author: James Hazelden, Eric Shea-Brown
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.DS, math.OC, q-bio.NC
arXiv:2605.12763v2 Announce Type: replace Abstract: Rich learning in recurrent neural networks often proceeds through sudden transitions in latent dynamics, but there is little theory predicting how gradient descent behaves during these events. We study the local learning geometry near codimension-o...
174. FishBack: Pullback Fisher Geometry for Optimal Activation Steering in Transformers ​
Author: Sihan Wang, Jiayi Zhao, Qingyan Cao, Hongbo Yao, Lin Shu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2605.17231v2 Announce Type: replace Abstract: Activation steering has emerged as a lightweight approach for modifying language model behavior without parameter updates, yet existing methods remain brittle: unstable across layers and prone to disturbing behavior unrelated to the target concept....
175. ImplicitTerrainV2: Wavelet-Guided Spatially Adaptive Neural Terrain Representation ​
Author: Haoan Feng, Xin Xu, Leila De Floriani
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.22556v2 Announce Type: replace Abstract: Digital elevation models (DEMs) underpin terrain analysis in Geographic Information Systems (GIS), but commonly as raster representation, they rely on interpolation for off-grid sampling and finite-difference operators for derivative-based analysis...
176. Open datasets and machine learning for two-phase heat transfer: a review following a spatial-temporal taxonomy ​
Author: Christy Dunlap, Ridwan Olabiyi, Firas Al-Hindawi, Hari Pandey, Stephen Pierson, Daniel Curl, Braden Stevens, Mohammad Ishraq Hossain, Annapurna Parjuli, Chinmaya Joshi, Ashif Iquebal, Han Hu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn
arXiv:2605.23037v2 Announce Type: replace Abstract: Two-phase heat transfer underpins boiling, condensation, immersion cooling, flow boiling, energy conversion, and electronics thermal management, but its coupled interfacial physics make data reuse and model comparison difficult. This narrative revi...
177. Nonlinear Data Integration via Kernel Methods for Data Collaboration Analysis ​
Author: Yamato Suetake, Yuta Kawakami, Shunnosuke Ikeda, Yuichi Takano
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2605.27219v2 Announce Type: replace Abstract: Collaborative analysis of decentralized confidential datasets is important, but direct sharing of original datasets is often restricted by privacy and institutional constraints. Data collaboration (DC) analysis transforms each dataset into privacy-...
178. TabCausal: Pretraining Across Causal Environments for Tabular Causal Discovery ​
Author: Zi-Rong Li, Si-Yang Liu, Tian-Zuo Wang, Han-Jia Ye
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.31156v2 Announce Type: replace Abstract: Causal discovery aims to recover directed causal relations from observational and interventional data, providing a basis for mechanistic understanding and reliable decision-making. Causal discovery foundation models (CDFMs) seek to amortize this pr...
179. BRo-JEPA: Learning Modular Transformations in Latent Space ​
Author: Divyansh Jha, Yuanfang Xie, Brennen Yu, Varan Mehra
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2606.01372v2 Announce Type: replace Abstract: Can neural networks learn algebraic rules from visual inputs, or do they merely fit observed patterns? We study this question using MNIST (or EMNIST letters) as states and modular arithmetic operations as actions in a JEPA-style world model. Standa...
180. Mos-Gen: A Generative Molecular Framework for Mosquito Insecticide Design ​
Author: Lina Wang, Yaning Cui, Zhifeng Gao, Ping Xing, Biao Jiang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.01846v2 Announce Type: replace Abstract: Mosquito-borne infectious diseases cause more than 700000 deaths worldwide each year. The long-term use of conventional chemical insecticides has induced serious resistance problems, creating an urgent need to develop novel, highly effective, and e...
181. The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics ​
Author: Pietro Barbiero, Giovanni De Felice, Mateo Espinosa Zarlenga, Francesco Giannini, Filippo Bonchi, Mateja Jamnik, Giuseppe Marra, Ruggero Noris
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2606.12289v2 Announce Type: replace Abstract: As Artificial Intelligence models grow in complexity, interpretability has become an indispensable tool for understanding, debugging, and controlling their computations. However, interpretability lacks general theories to deductively design interpr...
182. How Transparent is DiffusionGemma? ​
Author: Joshua Engels, Callum McDougall, Bilal Chughtai, Janos Kramar, Senthooran Rajamanoharan, Cindy Wu, Arthur Conmy, Asic Q Chen, Jean Tarbouriech, Min Ma, Brendan O'Donoghue, Jo~ao Gabriel Lopes de Oliveira, Rohin Shah, Neel Nanda
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.20560v2 Announce Type: replace Abstract: LLM reasoning transparency is a critical affordance for understanding model decisions, mitigating misuse and misalignment, and debugging surprising model behaviors. However, DiffusionGemma performs a larger fraction of its computation in a continuo...
183. Anti-Collapse Dynamics and the Emergence of Multi-Time-Scale Learning in Recurrent Neural Networks ​
Author: Lorenzo Livi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, physics.data-an
arXiv:2606.29519v3 Announce Type: replace Abstract: Long-range learning is hard for recurrent networks trained with stochastic gradient descent, because the influence of a past input fades with the lag $\ell$, and if it fades too fast the dependence cannot be learned from finite data. This fade is c...
184. Low-dimensional topology of deep neural networks ​
Author: Junyu Ren, Lek-Heng Lim
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.GT
arXiv:2606.31856v2 Announce Type: replace Abstract: We study layered models, including feedforward networks, ResNets, and transformers, by limiting each layer to a width of $d = 3$, i.e., $\mathbb{R}^3$ as representation space. This allows us to track how a neural network changes low-dimensional top...
185. SEE: Structure-aware Exploring & Exploiting for Long-horizon GUI Agent Trajectory Synthesis ​
Author: Zhuohang Fan, Beichen Zhang, Yuanfa Li, Changqiao Wu, Wei Liu, Jian Luan, Weigang Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18046v2 Announce Type: replace Abstract: Graphical User Interface (GUI) agents powered by vision-language models hold promise for automating real-world mobile tasks. However, progress is limited by the lack of high-coverage, long-horizon interaction trajectories collected from element-ric...
186. GEqTrain: A Configuration-Driven Framework for Retargeting Equivariant Graph Neural Networks Across 3D Scientific Tasks ​
Author: Daniele Angioletti, Marco Nobile, Vittorio Limongelli
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM
arXiv:2607.19083v3 Announce Type: replace Abstract: Equivariant graph neural networks provide a powerful modeling language for three-dimensional scientific data, but their reuse is often limited by implementations tied to specific tasks, outputs, and training regimes. We present GEqTrain, a configur...
187. ClockRoPE: Random Fourier Rotations for Temporal Routine Modeling ​
Author: Yiwen Chen, Joshua Ainslie, Krzysztof Choromanski, Xiang Gao, Su-Lin Wu, Yiping Yuan, Qian Sun
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.26369v2 Announce Type: replace Abstract: Rotary Position Embedding (RoPE) has been widely adopted in transformer-based large language models. However, its log-linear frequency schedule, originally designed to produce long-term attention decay, limits its adoption in domains with more comp...
188. A Physics-Informed Hybrid Neural Operator for Transient Magnetization Prediction in Power Magnetics ​
Author: Yachao Zhu, Qiujie Huang, Sinan Li, Yang Li, Gang Lei, Jianguo Zhu
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.SY, eess.SY
arXiv:2608.02965v3 Announce Type: replace Abstract: Magnetic components in high-frequency, high-power-density converters are increasingly driven by non-sinusoidal flux-density waveforms with fast transitions, minor-loop operation, dc bias, and temperature variation. Under these conditions, steady-st...
189. A Comparative Study of Feature Selection Methods for EHR Diagnosis Codes in Opioid Use Disorder Prediction ​
Author: Zihan Ding, Yinan Liu, Tengfei Ma, Rachel Wong, Xia Zhao, Richard N. Rosenthal, Fusheng Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.04180v2 Announce Type: replace Abstract: Feature selection is a critical step in electronic health record (EHR)-based predictive modeling, where input variables are often high-dimensional, sparse, noisy, and redundant. Large feature sets not only increase computational burden and overfitt...
190. Online Learning of Scale Parameters in Score-Driven Filters ​
Author: Fabrizio Lillo, Giulia Livieri, Gianluca Palmari
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ME, stat.ML, stat.TH
arXiv:2608.09218v2 Announce Type: replace Abstract: Score-driven filters update a time-varying parameter by multiplying a scaled log-likelihood score by a scale parameter that controls the magnitude of the update. We name this scale parameter gain, consider it a decision variable, and study its onli...
191. Geometric and Behavioral Stratification in Transformer Residual Streams ​
Author: Nelson Guda
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.12447v2 Announce Type: replace Abstract: Trained transformer models develop privileged bases: coordinate axes whose statistics differ from the rest of the residual stream. But what kind of direction does such a basis select? We investigate the prediction direction, the unembedding directi...
192. A Contract-Grade Verifier for LLM-Generated GPU Kernels, and a Native Blackwell Backward for the Gated-Linear-Recurrence Family ​
Author: Rishi Shah, Rishav Shrestha
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.DC
arXiv:2608.12700v2 Announce Type: replace Abstract: Systems that generate GPU kernels with language models report high correctness rates. Those rates come from a single loose test: run the kernel on a few random inputs at one fixed shape and accept it if the output is close to a reference. A kernel ...
193. Federated Compositional Muon Optimizer for Matrix-Wise Models ​
Author: Wang Yan, Feihu Huang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2608.12710v2 Announce Type: replace Abstract: Muon, a more recently developed optimizer, is useful for matrix-wise models in AI areas. Although many works have studied Muon and its variants, these methods are still not particularly well-suited for hierarchical structured problems. To fill this...
194. CoMedBench: A Multi-Source Benchmark of Synthetic Medical Data Fidelity and Downstream Utility ​
Author: Akanta Das, Farhad Al-Amin Dipto, Mrinmoy Sarkar Anto, David Rehkopf, Ayin Vala, Tanmoy Sarkar Pias
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12805v2 Announce Type: replace Abstract: Access to clinical data is essential for developing reliable healthcare machine learning systems, but direct use of electronic health records is constrained by privacy regulation, institutional review, data-use agreements, and the risk of re-identi...
195. The Integer Alibi: Localizing Cross-Kernel Divergence in INT8-Quantized LLM Inference ​
Author: Teng-Ruei Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.13756v2 Announce Type: replace Abstract: Two GPU kernels implementing the same scaled INT8 GEMM interface are usually treated as interchangeable. We test that assumption: holding the checkpoint, prompts, hardware, inference engine, decoding, and quantization configuration fixed, we swap o...
196. CrevasseSeg: A Label-Efficient UAV Crevasse Segmentation Framework ​
Author: Steven Wallace, William D. Harcourt, Richard Hann, Aiden Durrant, Somayajulu Sripada, Georgios Leontidis
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15790v2 Announce Type: replace Abstract: Crevasse mapping from uncrewed aerial vehicle (UAV) imagery matters for glaciological research and for field safety in glaciated terrain. Yet, pixel-level annotation of glacier surfaces is costly and requires domain experts. We introduce CrevasseSe...
197. NICE: Scale-Stable Perturbations for Graph Neural Network Explanations via Noise Corruption ​
Author: Ziluowen Luo, Jun Yin, Ruochen Liu, Ming Cheng, Shirui Pan, Chengqi Zhang, Senzhang Wang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16038v2 Announce Type: replace Abstract: Post-hoc Graph Neural Network (GNN) explainers commonly follow a Perturb-Query paradigm, inferring the importance of graph elements based on queried predictions to perturbed inputs. However, such perturbations often introduce substantial distributi...
198. OceanDepths: A Global Dataset of Paired Subsurface and Surface Ocean Observations ​
Author: Simon Donike, Ruben Cartuyvels, Antonino Ian Ferola, Elisa Carli, Diego Fernandez Prieto, Marie-Helene Rio
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.16373v2 Announce Type: replace Abstract: Despite comprising over 70% of its surface, the world's oceans are critically underobserved compared to the land surface or the atmosphere. Understanding the global ocean requires jointly observing its surface and subsurface structure, yet no stand...
199. A Data-Efficient Analytical Prior Machine Learning Framework for Sound Reduction Frequency Prediction in Helmholtz Resonators ​
Author: Jiaming Li
Published: 8/20/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16873v2 Announce Type: replace Abstract: High-fidelity finite-element simulations can provide accurate numerical predictions for side-branch resonators, but large simulation datasets are expensive to generate and purely data-driven surrogates may become unreliable when simulation-labelled...
200. Deep Learning Based on Generative Adversarial and Convolutional Neural Networks for Financial Time Series Predictions ​
Author: Wilfredo Tovar
Published: 8/20/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2008.08041v3 Announce Type: replace-cross Abstract: In the big data era, deep learning and intelligent data mining technique solutions have been applied by researchers in various areas. Forecast and analysis of stock market data have represented an essential role in today's economy, and a sign...
201. The Authenticity Gap in Human Evaluation ​
Author: Kawin Ethayarajh, Dan Jurafsky
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2205.11930v3 Announce Type: replace-cross Abstract: Human ratings are the gold standard in NLG evaluation. The standard protocol is to collect ratings of generated text, average across annotators, and rank NLG systems by their average scores. However, little consideration has been given as to ...
202. Doubly robust nearest neighbors in factor models ​
Author: Raaz Dwivedi, Caleb Chin, Sabina Tomkins, Predrag Klasnja, Susan Murphy, Devavrat Shah
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2211.14297v5 Announce Type: replace-cross Abstract: We introduce and analyze an improved variant of nearest neighbors (NN) for estimation with missing data in latent factor models. We consider a matrix completion problem with missing data, where the $(i, t)$-th entry, when observed, is given b...
203. EquiPocket: an E(3)-Equivariant Geometric Graph Neural Network for Ligand Binding Site Prediction ​
Author: Yang Zhang, Zhewei Wei, Ye Yuan, Chongxuan Li, Wenbing Huang
Published: 8/20/2026, 4:00:00 AM
Categories: q-bio.BM, cs.LG
arXiv:2302.12177v5 Announce Type: replace-cross Abstract: Predicting the binding sites of target proteins plays a fundamental role in drug discovery. Most existing deep-learning methods consider a protein as a 3D image by spatially clustering its atoms into voxels and then feed the voxelized protein...
204. Comprehensive framework for evaluation of deep neural networks in detection and quantification of lymphoma from PET/CT images: clinical insights, pitfalls, and observer agreement analyses ​
Author: Shadab Ahamed, Yixi Xu, Sara Kurkowska, Claire Gowdy, Joo H. O, Ingrid Bloise, Don Wilson, Patrick Martineau, Fran\c{c}ois B'enard, Fereshteh Yousefirizi, Rahul Dodhia, Juan M. Lavista, William B. Weeks, Carlos F. Uribe, Arman Rahmim
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2311.09614v5 Announce Type: replace-cross Abstract: This study addresses critical gaps in automated lymphoma segmentation from PET/CT images, focusing on issues often overlooked in existing literature. While deep learning has been applied for lymphoma lesion segmentation, few studies incorpora...
205. Spikformer V2: Join the High Accuracy Club on ImageNet with an SNN Ticket ​
Author: Zhaokun Zhou, Yijie Lu, Kaiwei Che, Wei Fang, Keyu Tian, Qihao Peng, Yuesheng Zhu, Shuicheng Yan, Yonghong Tian, Li Yuan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.NE, cs.CV, cs.LG
arXiv:2401.02020v2 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs), known for their biologically plausible architecture, face the challenge of limited performance. The self-attention mechanism, which is the cornerstone of the high-performance Transformer and also a biologically...
206. Predicting Male Domestic Violence Using Explainable Ensemble Learning and Exploratory Data Analysis ​
Author: Md Abrar Jahin, Saleh Akram Naife, Fatema Tuj Johora Lima, M. F. Mridha, Md. Jakir Hossen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CY, cs.LG
arXiv:2403.15594v4 Announce Type: replace-cross Abstract: Domestic violence is commonly viewed as a gendered issue that primarily affects women, which tends to leave male victims largely overlooked. This study presents a novel, data-driven analysis of male domestic violence (MDV) in Bangladesh, high...
207. On Stability in Optimistic Bilevel Optimization ​
Author: Johannes O. Royset
Published: 8/20/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.SY, eess.SY
arXiv:2408.13323v3 Announce Type: replace-cross Abstract: Solutions of bilevel optimization problems tend to suffer from instability under changes to problem data. In the optimistic setting, we construct a lifted formulation that exhibits desirable stability properties under mild assumptions that ne...
208. Efficient Dynamic Shielding for Parametric Safety Specifications ​
Author: Davide Corsi, Kaushik Mallik, Andoni Rodriguez, Cesar Sanchez
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO, cs.RO, cs.SY, eess.SY
arXiv:2505.22104v2 Announce Type: replace-cross Abstract: Shielding has emerged as a promising approach for ensuring safety of AI-controlled autonomous systems. The algorithmic goal is to compute a shield, which is a runtime safety enforcement tool that needs to monitor and intervene the AI controll...
209. On detection probabilities of link invariants ​
Author: Tuomas Kelom"aki, Abel Lacabanne, Daniel Tubbenhauer, Pedro Vaz, Victor L. Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: math.GT, cs.LG, math.QA
arXiv:2509.05574v3 Announce Type: replace-cross Abstract: We prove that, for many standard link invariants, both the proportion of distinct invariant values and the detection probability among prime alternating links with at most n crossings decay exponentially in n, with an explicit universal rate....
210. SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA ​
Author: Haozhou Xu, Dongxia Wu, Matteo Chinazzi, Ruijia Niu, Rose Yu, Yi-An Ma
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2509.25459v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) show promise in generating long-form scientific explanations that synthesize evidence and connect multiple factors. However, in long-form scientific question answering, LLMs often hallucinate, producing unsupporte...
211. Inverse Problems for Partial Differential Equations with Jump Discontinuities in Coefficients via Two-Stage Physics-Informed Deep Learning and Statistical Mixture Models ​
Author: Zhikun Zhang, Guanyu Pan, Xiangjun Wang, Yong Xu, Guangtao Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2510.14656v3 Announce Type: replace-cross Abstract: This work proposes a two-stage physics-informed deep learning framework that combines neural-network-based sampling with statistical inference and constrained parameter refinement. In the first stage, a dual-network physics-informed architect...
212. A multi-view contrastive learning framework for spatial embeddings in risk modelling ​
Author: Freek Holvoet, Christopher Blier-Wong, Katrien Antonio
Published: 8/20/2026, 4:00:00 AM
Categories: q-fin.RM, cs.LG
arXiv:2511.17954v2 Announce Type: replace-cross Abstract: Incorporating spatial information, particularly when related to climate, weather, and demographic factors, is crucial for improving underwriting precision and enhancing risk management in insurance. However, spatial data are often unstructure...
213. SparsePixels: Efficient Convolution for Sparse Data on FPGAs ​
Author: Ho Fung Tsoi, Dylan Rankin, Vladimir Loncar, Philip Harris
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, hep-ex
arXiv:2512.06208v4 Announce Type: replace-cross Abstract: Inference of standard convolutional neural networks (CNNs) on FPGAs often incurs high latency and a long initiation interval due to the deep nested loops required to densely convolve every input pixel regardless of its feature value. However,...
214. Large Language Models: A Mathematical Formulation ​
Author: Ricardo Baptista, Andrew Stuart, Son Tran
Published: 8/20/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, stat.ML
arXiv:2601.22170v2 Announce Type: replace-cross Abstract: Large language models (LLMs) process and predict sequences containing text to answer questions, and address tasks including document summarization, providing recommendations, writing software and solving quantitative problems. We provide a ma...
215. Fermi-Dirac thermal measurements: A framework for quantum hypothesis testing and semidefinite optimization ​
Author: Nana Liu, Mark M. Wilde
Published: 8/20/2026, 4:00:00 AM
Categories: quant-ph, cond-mat.stat-mech, cs.LG
arXiv:2603.04061v2 Announce Type: replace-cross Abstract: Quantum measurements are the means by which we recover messages encoded into quantum states. They are at the forefront of quantum hypothesis testing, wherein the goal is to perform an optimal measurement for arriving at a correct conclusion. ...
216. Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs? ​
Author: Jeonghye Kim, Xufang Luo, Minbeom Kim, Sangmook Lee, Dohyung Kim, Jiwon Jeon, Dongsheng Li, Yuqing Yang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2603.24472v4 Announce Type: replace-cross Abstract: Self-distillation has emerged as an effective post-training paradigm for LLMs, often improving performance while shortening reasoning traces. However, in mathematical reasoning, we find that it can reduce response length while degrading perfo...
217. From Diffusion to Flow: Efficient Motion Generation in MotionGPT3 ​
Author: Jaymin Bhan, JiHong Jeon, SangYeop Jeong
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2603.26747v3 Announce Type: replace-cross Abstract: Recent text-driven motion generation methods span both discrete token-based approaches and continuous-latent formulations. MotionGPT3 exemplifies the latter paradigm, combining a learned continuous motion latent space with a diffusion-based p...
218. Does Unification Come at a Cost? Uni-SafeBench: A Safety Benchmark for Unified Multimodal Large Models ​
Author: Zixiang Peng, Yongxiu Xu, Qin-Yi Zhang, Jiexun Shen, Yi-Fan Zhang, Hongbo Xu, Yubin Wang, Gaopeng Gou
Published: 8/20/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2604.00547v2 Announce Type: replace-cross Abstract: Unified Multimodal Large Models (UMLMs) integrate understanding and generation capabilities within a single architecture. While unified architectures expand multimodal capabilities, their safety implications remain important yet underexplored...
219. Attention Flows: Tracing LLM Conceptual Engagement via Story Summaries ​
Author: Rebecca M. M. Hicke, Sil Hamilton, David Mimno, Ross Deans Kristensen-McLachlan
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2604.06416v2 Announce Type: replace-cross Abstract: Although LLM context lengths have grown, there is evidence that their ability to integrate information across long-form texts has not kept pace. We evaluate one such understanding task: generating summaries of novels. When human authors of su...
220. SegWithU: Uncertainty as Perturbation Energy for Single-Forward-Pass Risk-Aware Medical Image Segmentation ​
Author: Tianhao Fu, Austin Wang, Charles Chen, Roby Aldave-Garza, Yucheng Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2604.15271v4 Announce Type: replace-cross Abstract: Reliable uncertainty estimation is critical for medical image segmentation, where automated contours feed downstream quantification and clinical decision support. Many strong uncertainty methods require repeated inference, while efficient sin...
221. FairNVT: Fair Classification via Noise Injection in Vision Transformers ​
Author: Qiaoyue Tang, Sepidehsadat Hosseini, Mengyao Zhai, Thibaut Durand, Greg Mori
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2604.16780v2 Announce Type: replace-cross Abstract: This paper presents FairNVT, a lightweight debiasing framework for pretrained transformer-based encoders that improves prediction fairness while preserving task performance. FairNVT is motivated by the intuition that reducing sensitive-attrib...
222. Convergent Evolution: How Different Language Models Learn Similar Number Representations ​
Author: Deqing Fu, Tianyi Zhou, Mikhail Belkin, Vatsal Sharan, Robin Jia
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2604.20817v2 Announce Type: replace-cross Abstract: Language models trained on natural text learn to represent numbers using periodic features with dominant periods at $T=2, 5, 10$. In this paper, we identify a two-tiered hierarchy of these features: while Transformers, Linear RNNs, LSTMs, and...
223. Nonlinear GENERIC-Embedded Neural Networks (N-GENNs): Learning GENERIC dynamics with non-quadratic dissipation potentials ​
Author: Vojt\v{e}ch Votruba, Zequn He, Weilun Qiu, Celia Reina, Michal Pavelka
Published: 8/20/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG
arXiv:2605.09058v2 Announce Type: replace-cross Abstract: We introduce Nonlinear GENERIC-Embedded Neural Networks (N-GENNs), a deep learning framework for discovering evolution equations of systems governed by the nonlinear GENERIC formalism (General Equation for Non-Equilibrium Reversible-Irreversi...
224. Adaptive AI Task Partitioning and Safe Offloading in Heterogeneous Edge-Cloud Continuum ​
Author: Akuen Akoi Deng, Eimantas Butkus, Alfreds Lapkovskis, Praveen Kumar Donta
Published: 8/20/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG, cs.NI, cs.PF
arXiv:2605.09623v2 Announce Type: replace-cross Abstract: In recent years, the use of artificial intelligence on resource-constrained IoT devices has grown significantly. However, existing approaches to AI task partitioning and offloading across the edge-cloud continuum typically rely on static meth...
225. Memory by Design: Probabilistic Sequence Layers ​
Author: Matthew Dowling, Hyungju Jeon, Cristina Savin, Il Memming Park
Published: 8/20/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2605.31163v3 Announce Type: replace-cross Abstract: We introduce the \emph{design-model framework}: a way to derive efficient recurrent sequence maps from explicit assumptions about memory. A design model writes evidence into memory by exact Bayesian filtering; a query- dependent readout produ...
226. SEAM: Shortcut-Aware Real-Time Detection of Scripted vs. Spontaneous Speech for Interview Guardrails ​
Author: Vsevolod (V.), Kovalev, Pranay Manocha
Published: 8/20/2026, 4:00:00 AM
Categories: eess.AS, cs.LG
arXiv:2606.06837v2 Announce Type: replace-cross Abstract: Scripted vs spontaneous speech detection is appealing for interview guardrails, but benchmark performance can be inflated by shortcuts tied to corpus identity, channel conditions, and recording artifacts rather than speaking style itself. We ...
227. Spectrally Safe Neural Operator Warm-Starts for Large-Scale Newton Solvers ​
Author: Jaemin Oh, Youngkyu Lee, Jerome Darbon, George Em Karniadakis
Published: 8/20/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA
arXiv:2606.21828v2 Announce Type: replace-cross Abstract: Neural operators are increasingly used to warm-start Newton solvers for nonlinear PDEs, on the premise that a low test error places the initial guess inside the basin of attraction. We show that this premise is unreliable. An operator trained...
228. Turning Off-Policy Tokens On-Policy: A Plug-in Approach for Improving LLM Alignment ​
Author: Yu Li, Xiuyu Li, Mingyang Yi, Jiaxing Wang, Liangxu Zhang, Zhaolong Xing, Zhen Chen
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.04728v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) post-training for large language models (LLMs) follows a efficient paradigm of "rollout then update", which inevitably results in off-policy training data. To resolve this, Importance sampling (IS) is proposed, whi...
229. CHM-Net: Center Heatmap-driven Macro-Micro Modeling Network for MRI-based Microbial Density Stratification ​
Author: Jiaming Liang, Haolin Chen, Tingting Li, Bowen Yu, Qianyan Long, Tinghe Zhang, Xi Zhong, Xiaowei Hu, Xiaoqi Sheng, Hongmin Cai
Published: 8/20/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2607.09812v2 Announce Type: replace-cross Abstract: Microbial density is clinically important for tumor assessment and treatment decision-making, and recent advances in deep learning suggest that it can be non-invasively inferred from multimodal MRI. In this work, MRI-based Microbial Density S...
230. From Adoption to Deployment: A Qualitative Study on AI Integration in Software Development Practice ​
Author: Mahzabin Tamanna, Elizabeth Lin, Sparsha Gowda, Laurie Williams, Dominik Wermke
Published: 8/20/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CR, cs.IR, cs.LG
arXiv:2607.16660v2 Announce Type: replace-cross Abstract: The increasing adoption of Large Language Models (LLMs) as AI components in modern software systems introduces distinct security risks to the software supply chain. While many considerations and safety mechanisms are in place for components o...
231. Constitutional Midtraining: Content Presence Drives Alignment Gains ​
Author: Desiree Cho, Cameron Tice, Bernie Hogan, Hunar Batra, Puria Radmard, Jun Zhao, Nigel Shadbolt
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG
arXiv:2607.26654v3 Announce Type: replace-cross Abstract: Post-training alignment is often shallow, eroding under fine-tuning. It remains untested as to whether constitutional midtraining interventions can produce durable alignment when cleanly isolated from post-training. We build a 394M-token cons...
232. Non-KKT Accumulation in Entropic Mirror Descent ​
Author: Kuangyu Ding, Kim-Chuan Toh
Published: 8/20/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DS
arXiv:2608.01658v3 Announce Type: replace-cross Abstract: For mirror descent generated by a Legendre kernel, perhaps one of the most basic question in optimization is this: must every accumulation point of a bounded mirror descent sequence be Karush--Kuhn--Tucker (KKT) stationary under proper stepsi...
233. Large-scale AI-Ready Data for Anti-Cancer Drug Response Modeling ​
Author: Vincent Lavelle, Yitan Zhu, Kaitlyn Marlor, Thomas Brettin, Rick Stevens
Published: 8/20/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG
arXiv:2608.11444v2 Announce Type: replace-cross Abstract: Drug response prediction (DRP) models are an active area of research in pharmacogenomics, with growing potential to accelerate the identification of effective anticancer drugs. However, their predictive performance is often constrained by lim...
234. Efficient Hessian-Free Methods for Multi-Objective Bilevel Optimization with Nonconvex Lower Level ​
Author: Yicong Jiang, Feihu Huang
Published: 8/20/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2608.12704v2 Announce Type: replace-cross Abstract: Multi-objective bilevel optimization has wide applications in the AI area such as automated learning and multi-task meta-learning. Although recently some works have been begun to study the multi-objective bilevel optimization, the proposed me...
235. Belayer: Efficient Fault Tolerance for LLM Agentic RL Training ​
Author: Jiecheng Zhou, Qinghao Hu, Peng Sun, Xingcheng Zhang, Weiming Zhang
Published: 8/20/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2608.14635v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents are increasingly trained with reinforcement learning in long-horizon, sandboxed environments. Unlike conventional RL, agentic RL couples GPU-intensive rollout engines with stateful environment containers whos...
236. RecurrentGPT: Expressive Depth through Recurrent Modulation in Transformers ​
Author: Amr Hegazy, Amr Alanwar, Mostafa Elhoushi
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.15062v2 Announce Type: replace-cross Abstract: Scaling transformer language models creates an inherent tension between expressivity and memory efficiency. While unique weights across layers preserve functional specialization---from input-grounding to abstract refinement---they incur a sub...
237. The Null Token Knows: Reducing Message-Free Hallucination in ASR and NMT ​
Author: Kirill Borodin, Vasiliy Kudryavtsev, Ivan Viakhirev, Grach Mkrtchian
Published: 8/20/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SD
arXiv:2608.15940v2 Announce Type: replace-cross Abstract: Modern encoder-decoder systems can produce fluent text even when their input contains no recoverable message. We study this failure in ASR and NMT through the models' reserved null tokens, asking whether the score for ending generation alread...