arXiv cs.LG - 2026-08-28 ​
252 items collected.
1. SLM-Conditioned Hierarchical Relation Routing for Labeled Property Graph Learning ​
Author: Michal Podstawski
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.26132v1 Announce Type: new Abstract: Labeled property graphs combine relational structure with heterogeneous textual and categorical properties attached to both nodes and relationships. Conventional graph neural networks typically represent these properties as static feature vectors, limi...
2. NeuronFuzz: Safety Neuron Guided Fuzzing for LLM Safety Evaluation ​
Author: Zhiyuan Xu, Muhammad Firhard Roslan, Joseph Gardiner, Sana Belguith, Lichao Wu
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR, cs.SE
arXiv:2608.26222v1 Announce Type: new Abstract: Safety evaluation is critical for assessing whether aligned Large Language Models (LLMs) remain robust against jailbreak attacks. Existing automated testing methods, however, largely rely on response-level feedback: each candidate prompt typically requ...
3. Pruning Binarized Neural Networks: A Dedicated Framework and Globally Weighted Algorithms ​
Author: Roan Rubiales, Jean Pierre David
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26233v1 Announce Type: new Abstract: Extreme compression of deep neural networks, up to full binarization, dramatically reduces memory footprint and arithmetic complexity, facilitating deployment on constrained edge hardware with field-programmable gate arrays (FPGAs) and microcontrollers...
4. Muon with Finite Newton-Schulz: The Smoothing Benefit in Nonsmooth Nonconvex Optimization ​
Author: Mingyi Li, Taira Tsuchiya
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML
arXiv:2608.26288v1 Announce Type: new Abstract: Muon has emerged as a strong optimizer for the matrix-valued parameters in large language model pretraining, approximately orthogonalizing its momentum with a few Newton-Schulz iterations. Existing theory either replaces this iteration with the exact p...
5. Algebraic Multigrid Acceleration for Efficient Label Spreading ​
Author: Antonia van Betteray, Jonathan Klees, Miriam Sch"afers, Matthias Rottmann
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26309v1 Announce Type: new Abstract: Modern machine learning models rely on large amounts of labeled data. However, manual annotation of large-scale datasets is expensive and time-consuming. Label spreading is a semi-supervised learning technique that addresses this challenge by propagati...
6. Privacy Without Regret: Differentially Private Inference-Time Alignment ​
Author: Ishi Jain, Nandini Bhattad, Sayak Ray Chowdhury
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26324v1 Announce Type: new Abstract: Best-of-N (BoN) sampling is the simplest and most widely deployed inference-time alignment strategy, but it suffers from two distinct problems: reward hacking, in which the selected response exploits errors in the proxy reward model, and the absence of...
7. Beyond Capability Benchmarks: Learning Operational Fingerprints of LLM Cloud Services from Production Incident Metadata ​
Author: Meiwei Zhang, Eduardo Miranda, Bruce Baynes, Suvigya Jain, Wanlong Chen, Tao He, Sergey Borodavkin
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26332v1 Announce Type: new Abstract: Managed LLM services are now part of real production systems, but model selection and service planning still rely heavily on capability benchmarks that reveal little about operational behavior after deployment. We present Operational Embedding (OpEmbed...
8. CG4AI: A Column Generation Framework for Training AI Models Under Constraints ​
Author: Youcef Magnouche, Abderrahmane Driouch, S'ebastien Martin, Pierre Bauguion
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DM
arXiv:2608.26375v1 Announce Type: new Abstract: Standard machine-learning training minimizes a loss function over a dataset, but does not guarantee that the resulting model will satisfy predefined rules or constraints on its outputs. In many real-world applications, ranging from autonomous systems t...
9. The Latent Diagnostic Taxonomy: A Framework for Constructing Classifiers and Diagnosing Their Decisions, Applied to Prompt Injection Detection ​
Author: Jaturong Kongmanee, Smile Thanapattheerakul
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CR
arXiv:2608.26423v1 Announce Type: new Abstract: This paper proposes a framework for constructing a classifier as a safeguard layer, and for developing a complementary diagnostic that identifies which of the classifier's confident decisions can be trusted. This framework, the Latent Diagnostic Taxono...
10. FedCMAPSS: A Benchmark for Federated Learning in Remaining Useful Life Estimation ​
Author: Amelia Sorrenti, Matteo Pennisi, Concetto Spampinato, Simone Palazzo
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26433v1 Announce Type: new Abstract: Data-driven prognostics and health management has emerged as a key enabler for Industry 4.0, yet the development of robust remaining useful life (RUL) estimation models is often limited by the scarcity of run-to-failure data. While federated learning o...
11. NeoTriFuse: Reliability-Aware Multimodal Fusion under Missingness Heterogeneity for Neonatal Mortality Risk Prediction ​
Author: Jiyuan Tian, Qincheng Shen, Ye Lin, Yu Gao, Haohui Lu
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26436v1 Announce Type: new Abstract: Neonatal mortality risk prediction from bedside monitoring data remains challenging due to extreme class imbalance, heterogeneous clinical risk factors, multi-scale temporal dynamics, and substantial missingness. We propose NeoTriFuse, a reliability-aw...
12. Subgraph Filtering for Fair Graph Neural Networks ​
Author: Haohui Lu, jiyuan Tian, Fangyu Zhou, Shahadat Uddin
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26437v1 Announce Type: new Abstract: Graph neural networks (GNNs) can exhibit unfair behavior even when sensitive attributes are excluded from node features, because graph topology and message passing propagate group-correlated signals under sensitive homophily. Existing fairness-aware GN...
13. Toward Equitable Low-Carbon Mobility: Fairness-Aware Demand Prediction for Expanding Bike-Sharing Systems ​
Author: Man Luo, Yixuan Zhao
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26451v1 Announce Type: new Abstract: Bike-sharing systems are an important component of low-carbon urban mobility, but continued expansion creates challenges in both cold-start prediction and equitable resource allocation. Newly deployed stations lack historical ridership records, causing...
14. Distributed Training using an Intelligent Network ​
Author: Nihar Shah, Ben Blier
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.NI
arXiv:2608.26453v1 Announce Type: new Abstract: Distributed training across a wide area network (WAN) is challenging, as continuous parameter exchange by islands of compute is constrained by limited bandwidth, high latency, and uneven topology. We propose making the network an active participant in ...
15. Diff Mining: Logit Differences Reveal Finetuning Objectives ​
Author: Greg Kocher, Robert West, Cl'ement Dumas, Julian Minder
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.26462v1 Announce Type: new Abstract: Finetuning has become the gold standard for refining existing behaviors and inducing new ones in language models, yet it often remains unclear exactly which behaviors emerge during this process. As models grow ever more capable, understanding finetunin...
16. Active Curriculum Refinement for Reinforcement Learning ​
Author: Zhenya Liu, Yuxin Chen
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26469v1 Announce Type: new Abstract: In many reinforcement learning (RL) domains, environments are connected by prerequisite relations, such as difficulty-increasing edits or parameter increments, which induce a directed acyclic curriculum graph (DAG). Although this structure is often exp...
17. Shared Actors Need Not Share Critics: Effects of Value Mismatch in Parallel Reinforcement Learning ​
Author: Zhenya Liu, Yang Meng, Zhuokai Zhao, Xuefeng Liu, Yuxin Chen
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26481v1 Announce Type: new Abstract: When a single policy is trained in parallel across multiple environments of the same task, such as procedurally generated levels, randomized dynamics, or curricula, implementations commonly use one critic across all sampled environments. Yet different ...
18. Bayesian methods and Markov chain Monte Carlo algorithms for curve reconstruction and point cloud data analysis ​
Author: Asir Intesar Tushar, Ioannis Sgouralis
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.26490v1 Announce Type: new Abstract: Point-cloud data routinely captured by modern imaging and sensor technologies provide detailed geometric descriptions of objects and environments, but their analysis is hindered by large data volumes, localization noise, and missing information. In add...
19. A Unified Framework for Fair and Personalized Decentralized Learning under Communication Constraints ​
Author: Krishnendu S. Tharakan, Carlo Fischione
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26493v1 Announce Type: new Abstract: Decentralized learning systems aim to collaboratively train models across multiple clients without relying on a central coordinator. While decentralization improves scalability, privacy, and robustness, it also exacerbates three fundamental challenges:...
20. A Single Suffix to Break Them All: Basin-Aware Jailbreaks for Merged Model Families ​
Author: Yu Zhe, Yixin Tan, Junhao Wei, Wang Chen
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.26506v1 Announce Type: new Abstract: Model merging enables combining multiple fine-tuned models without additional training, but its safety implications remain poorly understood. Prior work primarily attributes merging risks to unsafe constituent models, implicitly assuming that merging i...
21. Algorithmic Principles For Multiclass Learning Are Hard To Come By: Limits of Regularization and Proper Learning ​
Author: Julian Asilis, Shaddin Dughmi, Vatsal Sharan, Alec Sun, Shang-Hua Teng, Chang Wang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.26516v1 Announce Type: new Abstract: Two of the most fundamental questions in statistical learning theory are the following: which prediction problems are learnable, and how should they be learned? For the former, elegant answers often take the form of combinatorial dimensions. The latter...
22. High Probability Derivative Bounds for Random tanh Neural Networks on a Hypercube ​
Author: Josef Dick, Michael Feischl, Fabian Zehetgruber
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.26526v1 Announce Type: new Abstract: We establish high-probability bounds for mixed input derivatives of wide random neural networks whose activation derivatives satisfy a factorial growth bound. Our main result specializes these estimates to $\tanh$ networks with Xavier initialization. A...
23. Predicting Quantifiability from Primary Screens to Prioritize Dose-Response Profiling ​
Author: Sean Lim
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM
arXiv:2608.26538v1 Announce Type: new Abstract: High-throughput drug screening relies on low-cost primary assays to prioritize compounds for more expensive dose-response profiling, where potency is ultimately quantified. Current screening strategies largely focus on identifying compounds that will c...
24. Chart2SVG: Editable SVG Generation from Raster Chart Images ​
Author: Jinning Cui, Lu Chen, Haoyan Shi, Yue He, Chenglong Wang, Mengyu Zhou, Weidong Huang, Yunhai Wang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26544v1 Announce Type: new Abstract: We present Chart2SVG, a multimodal large language model that converts static raster charts into structurally organized, semantically enriched SVGs that support programmatic editing. By incorporating chart-specific semantic tokens into a vision-language...
25. Arrive and Survive: Scaling Safe Goal-Conditioned Policy Learning from One-Bit Failure Signals ​
Author: Guopeng Li, Yiyang Duan, Yiru Jiao, Chengcheng Xu
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2608.26571v1 Announce Type: new Abstract: Contrastive reinforcement learning (CRL) scales effectively in goal-conditioned tasks by casting policy learning into a self-supervised contrastive objective. However, in a failure-terminated Markov decision process, established CRL considers pre-failu...
26. Activation Outliers Matter: Robust Recovery for Quantized Multimodal LLMs ​
Author: Tanzila Rahman, Mehran Taghian Jazi, Yunke Peng, Zhuang Ma, Anandharaju Durai Raju, Yao Wang, Xing Huang, Hei Yi Mak, Shadan Golestan, Hoang Le, Yonghan Dong, Wei Guo, Yaoyuan Wang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26581v1 Announce Type: new Abstract: Low-bit quantization offers a promising avenue for reducing the computational and memory demands of Multimodal Large Language Models (MLLMs). Recent hardware support for low-precision formats, ranging from MXFP8 to ultra-low-bit formats such as MXFP4 a...
27. J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data ​
Author: Gyouk Chu, Myeongho Jeon, Eunho Yang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.26582v1 Announce Type: new Abstract: Self-evolving language models have recently emerged as a promising path toward superintelligence, with the advantage of reducing the cost of human supervision. While considerable progress has been made in verifiable domains, self-evolution in unverifia...
28. GRAS: Guided Reduced-Variance Proposals and Adaptive Selection for Training-Free Reward Alignment in Discrete Diffusion ​
Author: Kwanyoung Kim
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, q-bio.QM, stat.ML
arXiv:2608.26585v1 Announce Type: new Abstract: Discrete diffusion models have become a strong, widely adopted class of generators for sequence data, and steering them toward a downstream reward at inference time, without any retraining, is increasingly important. Such training-free steering is done...
29. SimCast-S2S: An Efficient Generative Model for Subseasonal Precipitation Forecasting via Transfer Learning from Climate Simulations ​
Author: Hiep V. Dang, Antonios Mamalakis
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26594v1 Announce Type: new Abstract: Subseasonal-to-seasonal (S2S) precipitation forecasting has substantial financial and societal impact, yet remains challenging because of weak predictive signals, high associated uncertainty, and the computational cost of operational systems, which con...
30. Technical Comparative Benchmarking Study: Advanced AI Hybrid Methods for Renewable Energy Farm Optimization and Forecasting ​
Author: Majid Masoumi, Asghar Dashtiy, Mohammad Dehghan, Mina Rajabi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26613v1 Announce Type: new Abstract: This study provides a comprehensive benchmarking of conventional machine learning (ML), ensemble learning, deep neural networks, recurrent architectures, Transformers, graph based models, and hybrid ensemble deep learning approaches under complementary...
31. Robust Neural Stimulation Response Modeling Through Meta-Learning and Pretraining ​
Author: Matthew J Bryan, Daniel C Muir, Felix Schwock, Azadeh Yazdan-Shahmorad, Rajesh P N Rao
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26649v1 Announce Type: new Abstract: Objective: Model-based closed-loop neural stimulation holds promise for therapeutic applications ranging from Parkinson's disease to sensory restoration, but deployment has been limited by two obstacles: 1) forecasting models for predicting the consequ...
32. When Privacy Hurts Mergeability: Geometry-Aware Model Merging under Differential Privacy ​
Author: Jin Liu, Junkang Liu, Ning Xi, Yinbin Miao, Dawei Wei, Ke Cheng, Jianfeng Ma
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26655v1 Announce Type: new Abstract: Model merging promises to construct a single multi-task model from independently fine-tuned task models without accessing the original task data. This makes it attractive when task data cannot be centralized, but released task models may still leak pri...
33. Simple Actors and Deep Critics for Scalable Reinforcement Learning ​
Author: Guhyeon Kang, Jaehwi Lee, Minhae Kwon
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26659v1 Announce Type: new Abstract: Recent progress in offline reinforcement learning (RL) has been driven by expressive generative actors such as diffusion and flow-matching policies, which capture multimodal behavior in offline datasets. However, these actors require multiple denoising...
34. Neural Regression with Embeddings for Numerical Attribute Prediction in Knowledge Graphs ​
Author: Rupesh Sapkota, Louis Mozart Kamdem Teyou, Moshood Yekini, Caglar Demir, Axel-Cyrille Ngonga Ngomo
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26729v1 Announce Type: new Abstract: In recent years, transductive knowledge graph embedding models have been applied to tasks such as link prediction and query answering. Although knowledge graphs often contain rich numerical attributes, most embedding models neglect them, limiting their...
35. Rethinking Message Passing as Retrieval for Text-Attributed Graph Learning ​
Author: Jintang Li, Yuhong Chen, Ruofan Wu, Binli Luo, Jiayi Ji, Hui Li, Rongrong Ji
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.26732v1 Announce Type: new Abstract: Graph neural networks (GNNs) are typically conceptualized as message-passing neural networks, yet it remains unclear why neighborhood aggregation reliably outperforms node-wise multilayer perceptrons (MLPs). Despite its empirical success, this paradigm...
36. Self-Augmented Diffusion Guidance for Physics-Informed Generation ​
Author: Akira Osaka, Naoya Takeishi, Takehisa Yairi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26748v1 Announce Type: new Abstract: Diffusion models can be used to generate spatiotemporal signals of physical phenomena, such as time-series images of fluid dynamics. However, a major limitation of standard diffusion models is that they do not incorporate constraints derived from the u...
37. Safety by Design: Realized-Cost Constraints for Contextual Bandits with Continuous Actions ​
Author: Spyros Dragazis, Aldo Pacchiano
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26755v1 Announce Type: new Abstract: Contextual bandits are a standard framework for sequential decision-making under uncertainty, with applications in clinical trials, dosage selection, recommendation systems, and autonomous systems. Safety is central in many of these applications, since...
38. Beyond Client Averaging: A Client-Independent Second-Order Stationary-Bias Component in Stochastic SCAFFOLD ​
Author: Yi-Ping Tang, Guan-Ju Peng
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.TH
arXiv:2608.26765v1 Announce Type: new Abstract: Existing constant-step analysis of stochastic \Scaf{} identifies a leading $O(\gamma/N)$ stationary mean bias and shows that higher-order bias can persist as the client count increases, but does not identify the first client-independent contribution at...
39. On the Indistinguishability of Human v/s AI Generated Text ​
Author: Jaee Ponde, Aritra Das, Mihir More, Debayan Gupta
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26797v1 Announce Type: new Abstract: The rapid improvement of LLMs has made distinguishing AI-generated text from human writing a pressing problem. This challenge is further amplified by paraphrasing tools designed to make machine-generated text appear more "human". We study how access to...
40. SAGE: Variate-Wise Semantic Augmentation for Vision-Language Time Series Forecasting ​
Author: Haizhao Fan, Xinyi Le
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.26829v1 Announce Type: new Abstract: Time series forecasting models operate on raw numerical sequences, lacking the semantic knowledge that domain experts implicitly leverage, such as the physical meaning of each variable, its statistical behavior, and its temporal dynamics. Recent effort...
41. Reinforcement Learning-Based Control of CAV Platoon Joining Maneuvers in Mixed Traffic ​
Author: Biao Yin, Abderrahmane Kasmi, Nadir Farhi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC
arXiv:2608.26860v1 Announce Type: new Abstract: Connected and automated vehicle (CAV) platooning offers a promising approach to improving road safety and traffic capacity. However, platoon control in real-world traffic is challenging due to uncertainty and heterogeneous driving behaviors. Reinforcem...
42. When Is the Sharp Covariance Envelope Tight? Feature-Only Geometry for Volume-Sampled Least Squares ​
Author: Kihun Rhee
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.26877v1 Announce Type: new Abstract: Prior analyses by Derezinski and Warmuth established all-size sampling identities, selected-OLS unbiasedness, and inverse moments for ordinary volume sampling, while their exact arbitrary-fixed-response loss and prediction-covariance formulas are at th...
43. Mitigating Strong-Modality Collapse in Multimodal Learning via Inverted Asymmetric Fusion ​
Author: Mary Ogbuka Kenneth, Foaad Khosmood, Abbas Edalat
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.MM
arXiv:2608.26879v1 Announce Type: new Abstract: Fusing multiple modalities is expected to improve model performance. However, on the MultiHuSE dataset, early, late, and symmetric attention fusion often fail to outperform the best unimodal baseline (text). Pathway isolation of a symmetric attention f...
44. A Layer Importance Metric for Quantization Accounting for the Speed-Quality Trade-off in Autoregressive Models ​
Author: Artem Safronov
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26926v1 Announce Type: new Abstract: Small language models (sLLMs) are nowadays hosted on devices with limited memory and computational budget. In an autoregressive setup, inference is memory-bandwidth bound: uniform quantization is often detrimental to such models, since their architectu...
45. Scaling Model-Generated Distillation Data Can Make Latent Teacher Traits More Recoverable ​
Author: Zhichen Dong, Zhixuan Liu, Yuyu Fan, Xiangtian Li, Shuyang Zhang, Chao Yang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.26958v1 Announce Type: new Abstract: Scaling model-generated data is usually viewed as improving distillation: more examples should increase coverage, reduce noise, and produce stronger students. We show a second effect: larger datasets can make subtle teacher-specific signals easier to d...
46. Gromov-Monge Flow Matching for Equivariant Graph Generation ​
Author: Moritz Piening, Christian Wald
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML
arXiv:2608.26961v1 Announce Type: new Abstract: Graphs are invariant under node permutations, motivating the use of permutation-equivariant architectures in generative models. In flow matching, however, symmetry may also enter the source--target coupling: once graph pairs are compared up to node rel...
47. Packora: Systematic Design for Generative Molecular Crystal Structure Prediction ​
Author: Nayoung Kim, Kiyoung Seong, Sungsoo Ahn
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci
arXiv:2608.26962v1 Announce Type: new Abstract: Molecular crystal structure prediction (CSP) is important in pharmaceuticals, agrochemicals, and organic electronics, where subtle differences in molecular conformation and packing can strongly affect material properties. We present Packora, a flow-bas...
48. Adversarial Training Without Input Gradients via Low-Rank Householder Expansions ​
Author: Tiana C. Johnson, Donsub Rim
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26963v1 Announce Type: new Abstract: This work concerns adversarial training against the small-norm adversarial examples that arise from the inherent input instability of a trained deep neural network. Examples in this class are small as measured in the relative $\ell^2$-norm, and therefo...
49. Graph-Based Pseudo-multimodal Contrastive Learning for 12-Lead ECG Representations ​
Author: Mengyu Wang, Kozo Okada, Takafumi Goto, Natsuko Jinba, Hiroki Yamaya, Kiyoshi Hibi, Tomoki Hamagami
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26964v1 Announce Type: new Abstract: 12-lead electrocardiogram (ECG) is a standard, non-invasive examination widely used for diagnosing coronary artery disease, where clinical interpretation relies on comparing waveform patterns across multiple leads. However, most existing ECG analysis m...
50. ClusterAttention: A training-free speedup of bidirectional attention ​
Author: Kasper Nordenram, Amelie Dittmann
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.26965v1 Announce Type: new Abstract: This paper introduces ClusterAttention, a general training-free speedup of bidirectional attention layers. Existing sparse attention methods either rely on structure in the input, such as order in language or spatial proximity in images, or use slow cl...
51. TEMPLAR Wales: A georeferenced environmental and toponymic dataset of Welsh settlements ​
Author: Oktay Karaku\c{s}, Can Eyupoglu
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.26970v1 Announce Type: new Abstract: Place names provide persistent records of how landscapes have been described and organised, but their quantitative reuse requires explicit separation between mapped places, lexical annotations and environmental measurements. TEMPLAR Wales is a georefer...
52. Terrain signatures in Welsh settlement names ​
Author: Oktay Karaku\c{s}, Can Eyupoglu
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.26978v1 Announce Type: new Abstract: Landscapes are named, but whether names retain measurable environmental information beyond broad geographic structure is rarely tested. We analysed 3,757 Welsh settlements using a frozen, source-audited 24-element lexical framework, preregistered outco...
53. Decentralized Multitask Learning over Learned Task Graphs ​
Author: Zirui Wan, Stefan Vlaski
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SP, eess.SY
arXiv:2608.26989v1 Announce Type: new Abstract: This paper investigates decentralized multitask learning over networks when the underlying task relationships are unknown. While existing graph-regularized multitask frameworks typically assume a known structure, practical settings often require learni...
54. Benchmarking_Fast_Domain_Adaptation_for_Unsupervised_Speech_Units ​
Author: Robin San Roman, Manel Khentout, Tu Anh Nguyen, Paul Michel, Yossi Adi, Emmanuel Dupoux
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26992v1 Announce Type: new Abstract: Representation learning has attracted great atten- tion and managed to reach good performances as a pretraining method for downstream tasks or as a first step towards unsu- pervised speech modeling. Yet, little is known about how such methods deal with...
55. Disentangling Optimization Scale from Preference Scale in DPO ​
Author: Ivan Kruzhilov
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27032v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is a widely used objective for aligning language models from preference data, with the coefficient $\beta$ commonly interpreted as controlling the KL constraint to a reference policy. We show that $\beta$ entangles ...
56. Performance Foundations of Parallel & Distributed Reasoning Language Models ​
Author: Maciej Besta, Leonard Schmidt, Lara Nonino, Robert Gerstenberger, Pierre Pang, Patrik Okanovic, Ales Kubicek, Tiancheng Chen, Baraq Lipshitz, Torsten Hoefler
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC, cs.PF
arXiv:2608.27046v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) and other RL-style post-training paradigms have been used for aligning large language models (LLMs) with reasoning standards. The resulting recent Reasoning Language Models (RLMs) such as DeepSeek-R...
57. Soft Active Electromyography Interface for Machine Learning-Enabled Silent Speech Recognition ​
Author: Yuta Kurotaki, Shusuke Yamakoshi, Reitaro Yoshida, Yutaka Isoda, Tamami Takano, Yuji Isano, Yusuke Miyake, Kentaro Kuribayashi, Hiroki Ota
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, cs.SD
arXiv:2608.27048v1 Announce Type: new Abstract: Silent speech recognition (SSR) provides an alternative communication pathway in the absence of audible speech. However, conventional approaches are limited by the need for constant facial attachment, privacy concerns, and unstable signal acquisition. ...
58. Unifying Detection and Adaptation in Task-Free Continual Learning ​
Author: Dezheng Han, Anbang Zhang, Zhihao Zhu, Shuaishuai Guo
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.27070v1 Announce Type: new Abstract: To mitigate catastrophic forgetting in downstream continual learning (CL) for large language models (LLMs), existing methods typically constrain parameter updates or introduce task-specific adaptation modules. However, these methods often rely on expli...
59. Emotional Preferences as Goal-Priority Regulation ​
Author: Shiqi Liu, Yihua Tan, Hu Fu, Guanyu Qi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.27072v1 Announce Type: new Abstract: A core question in decision-making for agents is whether the relative priorities of competing lower-level objectives can be determined by emotional preferences autonomously generated by higher-level goals, rather than being externally prespecified. Und...
60. Tabular Deep Learning for Algorithmic Trading: Cross-Regime Bayesian Optimisation for Equity Signal Generation ​
Author: Joshua Le Grice
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP, q-fin.TR
arXiv:2608.27076v1 Announce Type: new Abstract: Algorithmic trading now represents a market exceeding $20 billion, where even marginal gains in signal robustness can translate into economically significant returns. Existing evaluations of equity prediction models do not explicitly target regime robu...
61. Cone Extended Rayleigh Quotients for Directed Graph Learning: Minimax Spectral Certificates, Sensitivity, and Adaptive Control ​
Author: Yavdat Sh. Il'yasov, Nur F. Valeev
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27122v1 Announce Type: new Abstract: Directed graph learning naturally leads to trainable nonsymmetric propagation operators with distinct right and left spectral structures. Building on the two-sided cone Rayleigh framework for generalized pencils [ B_\theta-\lambda G, ] we develop a l...
62. TRACE-CRC: Trajectory-Adaptive Conformal Risk Control for Multi-Step Channel State Information Prediction ​
Author: Kiarash Rezaei, Mehdi Sattari, Javad Aliakbari, Tommy Svensson, Paolo Monti, Carlos Natalino
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.27124v1 Announce Type: new Abstract: Reliable prediction of time-varying channel state information (CSI) is essential for efficient wireless communication. Each CSI frame is a matrix-valued representation of the wireless channel response, and a sequence of CSI frames forms a temporal chan...
63. Ultra Low-Power, Lightweight, Probabilistic RSS-Based Path Reconstruction: A System for Landscape-Scale Bee Tracking ​
Author: Christopher J. Noroozi, Joseph L. Woodgate, Michael Mangan, Michael T. Smith
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27152v1 Announce Type: new Abstract: Applications in fields such as movement ecology, Internet of Things or robotics share the need for systems that localize devices that are too small and power constrained to implement GNSS (Global Navigation Satellite Systems). Alternative low-power loc...
64. Inductive Correlation Clustering with Graph Neural Networks ​
Author: Francesco Paolo Nerini, Francesco Bonchi, Arijit Khan, Andr'e Panisson
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.DS
arXiv:2608.27153v1 Announce Type: new Abstract: Correlation Clustering (CC) is a natural formulation of clustering in combinatorial optimization, which uses a graph representation of the input and does not require a pre-specified number of clusters. Given $n$ objects and a pairwise similarity functi...
65. Diffusion Policies for Short-Horizon Planning in Robot Crowd Navigation ​
Author: Wendong Li, Jochen Garcke
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27158v1 Announce Type: new Abstract: Robot crowd navigation requires safe and efficient decision-making under dense, dynamic, and multimodal human--robot interactions. Existing reinforcement-learning methods typically output a single reactive action at each timestep, which limits their ab...
66. TraceBench: Controlled Evaluation of LLM Agents for Time-Series Root-Cause Attribution ​
Author: Tommaso Bendinelli, Artur Dox, Christian Holz
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27182v1 Announce Type: new Abstract: LLM agents are increasingly applied to anomaly detection and root-cause analysis in time-series observations collected from real-world systems; however, their performance on these tasks has not been systematically evaluated under controlled conditions....
67. When Interference Graphs Evolve: Doubly Robust Estimation of Dynamic Peer Effects ​
Author: Xiaojing Du
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2608.27187v1 Announce Type: new Abstract: Peer effects are difficult to estimate when interaction graphs evolve because pre-assignment network history, dynamic peer exposure, and post-assignment network change have distinct causal roles. We introduce a controlled contrast framework that indexe...
68. Common Geodesics Do Not Guarantee Fisher Consistency of the Structured SVM: Minimal Counterexamples and a Tree-Metric Classification ​
Author: Jintao Fei, Jiangying Luo
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27203v1 Announce Type: new Abstract: A known necessary condition for Fisher consistency of the structured support vector machine requires the task loss to be a metric for which every output triple has a common geodesic point. We show that this condition is not sufficient for the canonical...
69. Profit based evaluation of machine learning for nitrogen recommendations in winter wheat ​
Author: Xulong Wang, Po Yang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27205v1 Announce Type: new Abstract: Nitrogen rates for winter wheat are set before the season, under unknown prices and weather. The standard UK advice does not respond to prices, yet recent price swings moved the most profitable rate by tens of kilograms per hectare. Machine learning is...
70. HALO: A Heterogeneity-Aware Language-Aligned IMU Foundation Model for Open-Set Human Activity Recognition ​
Author: Zihan Ding, Liyu Zhang, Xiaomin Ouyang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27233v1 Announce Type: new Abstract: Human Activity Recognition (HAR) using inertial measurement units (IMUs) enables a wide range of applications, yet the field still lacks a unified model that can generalize across diverse subjects, devices, and activities. Training such a model is diff...
71. Importance Scoring of Transformer Attention Heads in Learning Tabular Data ​
Author: Ahmad Jad Allah, Kazi F. Akhter, Md. Kamrozzaman Bhuiyan, Manar D. Samad
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27241v1 Announce Type: new Abstract: Computationally demanding and opaque deep learning models can be better understood and optimized by analyzing how they transform data. While deep transformers have been widely studied in computer vision and natural language processing, their applicatio...
72. Circuit Condensation: Post-Training that Concentrates a Behavior's Causal Circuit ​
Author: Sai Adith Senthil Kumar
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27254v1 Announce Type: new Abstract: One approach to mechanistic interpretability explains behavior through circuits: the components and connections that carry it. Frozen discovery often returns hundreds of edges, making them hard to inspect, compare, or verify exhaustively. We introduce ...
73. Making Latent Evolution Explicit: Operator-Structured Transitions for World Action Models ​
Author: Xiaoxiao Lu, Yunlong Dong, Jiahao Shi, Ye Yuan
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27259v1 Announce Type: new Abstract: World Action Models (WAMs) augment robot policies by predicting how task-relevant scene states may evolve under interaction. Recent WAMs increasingly perform such prediction in latent representation spaces, avoiding full appearance-level generation whi...
74. MM-Spectrum: Multimodal Multi-spectral Molecular Structural Elucidation with a Stable MoE Framework ​
Author: Hai-tao Yu, Nan Min, Zheng Fang, Hongyu Zhan, Yusen Tan, Yuhan Wang, Jun Xia
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27286v1 Announce Type: new Abstract: Inferring molecular structures from multimodal spectroscopic measurements requires integrating complementary yet highly heterogeneous signals. However, the common paradigm of directly concatenating multispectral sequences can exhibit anomalous performa...
75. QuantumBoostNet: A Hybrid Classical-Quantum Architecture for Enhanced Accuracy in Cardiac Ultrasound View Identification ​
Author: Mihai Udrescu-Milosav, Stefan-Alexandru Jura, Mihai Udrescu, Gerhard-Paul Diller
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27302v1 Announce Type: new Abstract: Accurate identification of the correct view or angle in cardiac ultrasound (echocardiogram) is a critical component of cardiologic imaging. This step is essential for precise anatomical interpretation, reliable measurement, and the reduction of clinica...
76. Beyond Parallel Blindness: Information Floors and Model Gaps in Block Drafting ​
Author: Xinwei Qiang, Xiang Fang, Chang Chen, Yue Guan, Yufei Ding
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IT, math.IT
arXiv:2608.27339v1 Announce Type: new Abstract: Block drafters propose several tokens in one forward pass, before earlier target tokens are realised. Their rejection mixes two losses: missing within-block path information and imperfect modelling of observable information. Accepted length cannot dist...
77. Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO ​
Author: Yunpeng Ba, Zhi Zheng, Yue Xie, Jiaqing Li, Xialiang Tong, Tao Zhong, Mingxuan Yuan, Zhichao Lu, Xuyang Wu, Zhenkun Wang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.27351v1 Announce Type: new Abstract: Evolution Strategies (ES) have recently emerged as a memory-efficient post-training paradigm for LLM reasoning. However, the optimization behavior of ES remains understudied, making it hard to define its advantage scope compared to mainstream post-trai...
78. A Dynamic Likelihood Approach to Filtering for Advection-Diffusion Dynamics ​
Author: Johannes Krotz, Juan M. Restrepo, Jorge Ramirez
Published: 8/28/2026, 4:00:00 AM
Categories: math.DS, cs.LG, math.ST, stat.TH
arXiv:2406.06837v2 Announce Type: cross Abstract: A Bayesian data assimilation scheme is formulated for advection-dominated advective and diffusive evolutionary problems, based upon the Dynamic Likelihood (DLF) approach to filtering. The DLF was developed specifically for hyperbolic problems -waves-...
79. Generative Monte Carlo Sampling for Constant-Cost Particle Transport ​
Author: Joseph A. Farmer, Aidan Murray, Johannes Krotz, Ryan G. McClarren
Published: 8/28/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG
arXiv:2512.13965v1 Announce Type: cross Abstract: We present Generative Monte Carlo (GMC), a novel paradigm for particle transport simulation that integrates generative artificial intelligence directly into the stochastic solution of the linear Boltzmann equation. By reformulating the cell-transmiss...
80. Recipes for Steering and Scaling LLMs via Sampling ​
Author: Jiajun He, Zongyu Guo, Jos'e Miguel Hern'andez-Lobato, Yuanqi Du
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.26120v1 Announce Type: cross Abstract: Large Language Models (LLMs) are probabilistic models, typically defined by an autoregressive factorization. While recent work has begun to study richer target distributions beyond the base model, the sampling strategies remain highly inefficient. In...
81. Can a Model Catch Its Own Hallucinations for Free?: Label-Free Doubt Signals Hold Their Own Against a Labelled Dataset for Abstention ​
Author: Ali Asaria, Tony Salomone, Deep Gandhi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.26121v1 Announce Type: cross Abstract: Large language models state false facts as fluently as true ones, yet a model often "knows" internally when it is on shaky ground: the probability it assigns to its own answer tends to dip on the facts it gets wrong. The usual way to act on this, tea...
82. Graph-Based Modeling of Financial Volatility Dynamics ​
Author: Chuanzhen Wang, Alice Zhang, Wei Chen, Michael Brown
Published: 8/28/2026, 4:00:00 AM
Categories: q-fin.ST, cs.CL, cs.LG
arXiv:2608.26127v1 Announce Type: cross Abstract: Accurate forecasting of realized volatility ($RV$) is crucial for risk management and derivatives pricing. Although the implied volatility ($IV$) surface offers rich informational content, prevailing methods that treat it as a static image fail to ca...
83. FIRSTPASS: A Multi-Domain, Multi-Round Peer Review Dataset Grounded in Real Editorial Outcomes ​
Author: Prabhjot Singh, Somnath Luitel, Manmeet Singh, Josh Durkee
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.26129v1 Announce Type: cross Abstract: Scientific peer review datasets have trained AI systems exclusively on Computer Science and Machine Learning venues, producing models that critique ablation studies yet have never seen a biology reviewer demand contamination controls or a chemist que...
84. Interpretable, Fairly Evaluated Automated L2 Speaking Assessment that Beats the Single-Human Ceiling and Why Pause Encoding Does Not Change LLM Fluency Scores ​
Author: Eichi Uehara
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SD
arXiv:2608.26137v1 Announce Type: cross Abstract: Second-language (L2) English learners can rarely rehearse speaking with a partner. Speaking is also the most anxiety-laden skill. These gaps drive a fast-growing market for automated speaking practice and scoring. But an automated score is trustworth...
85. Cross-Platform Generalisation Failure in Mental Health Natural Language Processing: A Five-Axis Fairness Audit of Transformer Models on Social Media ​
Author: Rajveer Singh Pall, Sameer Yadav
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.26138v1 Announce Type: cross Abstract: We introduce the Cross-Platform Fairness Evaluation (CPFE) framework -- a five-axis audit protocol covering discriminative performance, calibration, statistical significance, prediction equity, and attribution stability -- and apply it to four transf...
86. Affix Cache for Diffusion Large Language Models ​
Author: Kaihua Liang, An Zhong, Xin Tan, Zafar Ayyub Qazi, Hong Xu, Jian Weng, Marco Canini
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.26140v1 Announce Type: cross Abstract: Diffusion Large Language Models (DLLMs) enable non-autoregressive decoding and bidirectional context modeling, but efficient inference remains challenging. Unlike autoregressive systems, whose key-value (KV) cache can be reused for shared prefixes, D...
87. AdaThinking-E: One-Token Entropy Regulation for Adaptive Thinking ​
Author: Zining Wang, Tongkun Guan, Boming Chen, Zhentao Guo, Jianqiang Liu, Chao Jin, Chen Duan, Kai Zhou, Pengfei Yan, Wei Shen, Xiaokang Yang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.26141v1 Announce Type: cross Abstract: Multimodal large language models have demonstrated strong document reasoning capabilities by incorporating explicit thinking processes. While this capability significantly improves performance on challenging tasks, current models apply such deep reas...
88. Methodological and Conceptual Framework for 5D Multi-Table Analysis: A Unified Approach for Complex Data Reuse ​
Author: Edouard Lansiaux, Hugo Kazzi, Aur'elien Loison, Slim Hammadi, Emmanuel Chazard
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.26149v1 Announce Type: cross Abstract: Multi-table learning remains a major challenge in machine learning for healthcare and other complex information systems. Relational data combine several sources of complexity, including large data volume, high-dimensional variables, high-cardinality ...
89. Explainable Artificial Intelligence for Customer Churn Prediction in Telecommunications: A Framework for CRM Integration ​
Author: Sandeep Gaddamwar
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.26151v1 Announce Type: cross Abstract: Subscriber attrition is a costly, persistent challenge for telecommunications providers, with monthly churn of roughly 1.9% in mature markets eroding billions in revenue annually. Predictive models can flag at-risk customers accurately, yet they are ...
90. Selection Bias Correction in Retail Intelligence ​
Author: Spandan Ghose Chowdhury
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.26156v1 Announce Type: cross Abstract: Retail intelligence often relies on monitoring popular, high-velocity products, potentially biasing economic indicators by ignoring the "long tail" of niche items. This simulation study investigates selection bias in inflation estimation and compares...
91. Refusal Is Not Robustness: Auditing Confident Fabrication in Large Language Models on a Provably Uninformative Clinical Pain Speech Transcript ​
Author: Sagnik De, Sreenija Pavuluri
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.ET, cs.LG, eess.AS
arXiv:2608.26167v1 Announce Type: cross Abstract: Hallucination and abstention benchmarks rarely establish that a model could not have known the correct answer, making it difficult to distinguish appropriate abstention from an unsupported prediction. Seven large language models were evaluated on the...
92. ClassVision: AI-Powered Classroom Attendance System ​
Author: Ankit Kumar Aggarwal, Veerabhadra Rao Marellapudi, Ovadia Sutton, Youshan Zhang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CV, cs.LG
arXiv:2608.26173v1 Announce Type: cross Abstract: Students and working professionals have to go through the attendance process every day. Traditional methods of marking attendance using pen and paper or online platforms are human-intensive and time-consuming. To address the challenges in manual atte...
93. When the Canonical Completion Is Wrong: Formalizing and Measuring the Jump in Large Language Models ​
Author: Dai Shi, Xiaoyu Li, Jos'e Miguel Hern'andez-Lobato
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.LO
arXiv:2608.26187v1 Announce Type: cross Abstract: Whether large language models (LLMs) can perform the abductive leap from evidence to a new system of axioms, commonly referred to as a jump, has recently attracted considerable debate. A prominent position holds that LLMs are structurally incapable o...
94. Invocation-Level Reliability of Tool-Using Agents ​
Author: Afiya Noorain, Subhranshu Mohanty, Amritesh Banerjee, Abhijit Dasgupta
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.26189v1 Announce Type: cross Abstract: Tool-using agents fail two ways: choosing the wrong tool, or forming wrong arguments, and an early failure of either kind can silently corrupt everything downstream. We measure a correct-invocation rate that separates the two, under both a clean teac...
95. GameWAM: A World Action Model for Video Games ​
Author: Yuncheng Guo, Zhanqiu Zhang, Yiwen Guo, Weijia Li
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG
arXiv:2608.26200v1 Announce Type: cross Abstract: Modern video games combine first-person perception, rapid visual changes, persistent world state, and heterogeneous native controls. Existing game agents map visual and task context directly to actions but lack explicit world dynamics modeling, where...
96. Fairness Invariants: A Relational Approach to Explaining and Mitigating Fairness Bugs ​
Author: Ranit Debnath Akash, Ashish Kumar, Gang Tan, Saeid Tizpaz-Niari
Published: 8/28/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2608.26209v1 Announce Type: cross Abstract: Data-driven software systems are increasingly deployed in high-stakes socio-economic domains, from criminal justice to financial lending. However, these systems often exhibit individual discrimination---unjustified disparities in which a program yiel...
97. Real-time virtual circuits for plasma shape control via neural network emulators: integration and testing in the MAST-U PCS ​
Author: Matthew J. Marshall, Edward Jones, Graham J. McArdle, Alasdair Ross, Kamran Pentland, Nicola C. Amorisco, Charles Vincent, Martin Kochan, Colin Hogben, Graham Jones, Adam Stephen, George K. Holt, Adriano Agnello
Published: 8/28/2026, 4:00:00 AM
Categories: physics.plasm-ph, cs.LG
arXiv:2608.26216v1 Announce Type: cross Abstract: The deployment of advanced, AI-enabled control algorithms in tokamak experiments requires robust integration with existing plasma control system (PCS) architectures and extensive pre-experimental validation. In this contribution, we describe the inte...
98. TRACE: Retrospective Streaming Generation of Physical Fields under Sparse Structured Sensing ​
Author: Xinyu Zhang, Lihao Chen, Panqi Chen, Lei Cheng, Ting Zhang, Jianlong Li, Shikai Fang
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.26219v1 Announce Type: cross Abstract: Reconstructing continuous physical fields from sparse measurements is central to scientific monitoring, inverse modeling, and digital-twin construction. Generative reconstruction has recently emerged as a promising paradigm for this task by learning ...
99. Prompt Sensitivity of Generative Agents: Evidence from an Epidemic Model ​
Author: Ross Williams, Niyousha Hosseinichimeh
Published: 8/28/2026, 4:00:00 AM
Categories: physics.soc-ph, cs.AI, cs.LG, cs.MA
arXiv:2608.26221v1 Announce Type: cross Abstract: As generative AI gains traction, researchers are investigating its potential to serve as proxies for humans. From undergoing cognitive psychology experiments to experiencing an epidemic, generative agents, agents powered by generative AI models, prod...
100. Classical and Hybrid Quantum Machine Learning for Trigger-Like Event Selection on CMS Open Data: An Eight-Qubit, PCA-Constrained Benchmark ​
Author: Tariq Mahmood, Muhammad Awais Rafique, Talab Hussain, Juan Pablo Perez Aguilar, Alfredo Raya, Muhammad Ahsan
Published: 8/28/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-ex
arXiv:2608.26224v1 Announce Type: cross Abstract: Event triggering sits at the heart of high-energy physics, where the rare events of interest must be retained while an overwhelming background is discarded under tight latency and bandwidth budgets. This work compares four classical machine learning ...
101. A causal graph-informed temporal convolution architecture for interpretable retail electricity price forecasting ​
Author: Yufan Ji, Abdollah Shafieezadeh, Noah Dormady
Published: 8/28/2026, 4:00:00 AM
Categories: stat.AP, cs.LG
arXiv:2608.26234v1 Announce Type: cross Abstract: Retail electricity markets in deregulated systems face significant price volatility and complex interactions with forward and futures products, posing challenges for effective operational decision-making. This study introduces a Causal Graph-Informed...
102. Constraint-Aware Physics-Informed Neural Networks for Static Shape Estimation of Co-Manipulative Continuum Robots ​
Author: Rana Danesh, Pari Qarehdaghi, Farrokh Janabi-Sharifi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.26273v1 Announce Type: cross Abstract: Static shape estimation of co-manipulative continuum robots (CCRs) is challenging because the continuum arms and manipulated flexible object form a closed chain that must satisfy both static equilibrium and geometric loop-closure constraints. This pa...
103. Multi-Dataset Inverse Problem Solving with Distributed Generative AI ​
Author: Daniel Lersch, Steven Goldenberg, Johann Rudi, Markus Diefenthaler, Kevin Brager, Xingfu Wu, Yaohang Li, Nobuo Sato
Published: 8/28/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2608.26283v1 Announce Type: cross Abstract: Extracting a shared set of unknown, not directly measurable quantities from multiple, heterogeneous datasets is a common challenge across scientific domains. A prominent example is the combination of datasets obtained from different measurements with...
104. On Scope Classification and Current Knowledge-Editing Benchmarks: A Negative Result, with INLAY as a Gradient-Free Case Study ​
Author: Aditya Pratap Singh
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.26292v1 Announce Type: cross Abstract: Every memory-based knowledge editor in the SERAC lineage depends on a scope decision: given a query, does a stored edit apply? We report that current knowledge-editing benchmarks cannot measure this decision at all. Using INLAY, a gradient-free edito...
105. District-Level Food Environment Indicators and Social Vulnerability in S\~ao Paulo ​
Author: Pedro Lemes Sixel Lobo, Eric Tokuda, Kuruvilla Joseph Abraham, Roberto Fray, Dirce Maria Marchioni, Alexandre Cl'audio Botazzo Delbem, Rogerio Salvini
Published: 8/28/2026, 4:00:00 AM
Categories: physics.soc-ph, cs.LG, stat.AP
arXiv:2608.26299v1 Announce Type: cross Abstract: Urban food environments may reflect broader socioeconomic inequalities, but district-level evidence remains limited in Brazilian cities. This study examined whether indicators of food retail and street-market availability discriminate between levels ...
106. When Is Noise Response Universal? Tokenization as the Hidden Variable in Language Models ​
Author: Yefan Tao, Gerald Friedland, Luyang Kong
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.26319v1 Announce Type: cross Abstract: The performance of textual neural models often degrades when their inputs are corrupted by noise such as typos, OCR errors, or dropped words. We study the degradation rate across neural models, both sentence embeddings and decoder-only LLMs, and find...
107. Cross-simulator transfer with foundation model summaries: Towards robust SKA-era reionization inference ​
Author: Yannic Pietschke, Caroline Heneka, Ayodele Ore, Romain Meriot
Published: 8/28/2026, 4:00:00 AM
Categories: astro-ph.CO, astro-ph.IM, cs.LG, physics.data-an
arXiv:2608.26354v1 Announce Type: cross Abstract: Simulation-based inference (SBI) for parameter estimation is vulnerable to model misspecification: neural summaries and density estimators trained on a specific forward model typically fail when applied to data drawn from another model, or from real ...
108. Finding the Right Evidence: Factor-Guided Coarse-to-Fine Reasoning for Long Videos ​
Author: Baixuan Xu, Yinyui Xu, Tianshi Zheng, Zhaowei Wang, Weiqi Wang, Haochen Shi, Jiayu Liu, Qing Zong, Xiyu Ren, Xinyu Geng, Zhitao He, Yangqiu Song
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.26355v1 Announce Type: cross Abstract: While LVLMs rapidly improve, long-video question answering still remains challenging: relevant evidence is sparse, and question-relevant context often fails to provide cues that discriminate the correct answer from plausible alternatives. Diagnostic ...
109. Co-Evolving Structured Knowledge and Reasoning in Language Models ​
Author: Ryan Thomas Noonan, Linxi Zhao, Menghan Xu, Akanksha Sarkar, Mihir Mishra, Dongyoung Go, Kilian Q. Weinberger, Yoav Artzi, Jennifer J. Sun
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.26386v1 Announce Type: cross Abstract: Retrieval-augmented methods improve factual accuracy by grounding language models in external knowledge, but retrieving over unstructured text often introduces irrelevant context and offers limited control over the retrieved information. Structured k...
110. LowRankArena: A Standardized Evaluation Platform for SVD-Based LLM Compression ​
Author: Zishan Shao, Lixun Zhang, Kangning Cui, Wenhao Wu, Jinhee Kim, Yixiao Wang, Ting Jiang, Hancheng Ye, Qinsi Wang, Fan Yang, Danyang Zhuo, Yiran Chen, Hai Li
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.26389v1 Announce Type: cross Abstract: SVD-based low-rank compression has become a fast-growing direction for reducing the memory and computational cost of large language models (LLMs). However, meaningful comparison across existing studies remains difficult as prior evaluations use varie...
111. Towards a universal meta-optics solver via large language models ​
Author: Huanshu Zhang, Lei Kang, Yuyan Chen, Luxiang Wang, Zhaolong Cao, Douglas H. Werner
Published: 8/28/2026, 4:00:00 AM
Categories: physics.optics, cs.LG
arXiv:2608.26417v1 Announce Type: cross Abstract: Metasurface design increasingly requires fast models that can operate across structurally distinct device families, rather than retraining a separate surrogate for every geometry class. Conventional neural network surrogates often depend on fixed-dim...
112. Interpreting Latent Protein Language Model Features with Geometric Annotations ​
Author: Siddharth Setlur, Djordje Mihajlovic, Darrick Lee
Published: 8/28/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG
arXiv:2608.26419v1 Announce Type: cross Abstract: Protein language models (pLMs) encode information about protein sequences which enable downstream tasks such as structure prediction, but their internal representations are not well understood. Sparse autoencoders (SAEs) provide a promising tool to d...
113. Vowel Signs Are Not Letters: A Pre-tokenization Ceiling on Multilingual Tokenizer Fertility ​
Author: Sajal Regmi, Siddhartha Pudasaini, Chetan Phakami Pun
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.26449v1 Announce Type: cross Abstract: Byte-level BPE tokenizers that use the HuggingFace ByteLevel pre-tokenizer inherit GPT-2's word regex, where a word is defined as \p{L}+, one or more Unicode letters. In abugida scripts, vowels are written as combining marks; this pattern therefore s...
114. Systematic Literature Review of Machine Learning Models and Applications for Text Recognition ​
Author: Nuzhat Khan, Ab Al-Hadi Ab Rahman, Shahriyar Masud Rizvi, Ibrahim Yousef Alshareef, Muhammad Nadzir Marsono, Muhammad Paend Bakht, Mohd Shahrizal Rusli, Shahidatul Sadiah
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.26500v1 Announce Type: cross Abstract: Optical Character Recognition (OCR) for text recognition using machine vision has significantly improved, particularly when handling heterogeneous textual data. Traditional OCR models struggle with script variations, writing styles, and degraded docu...
115. Sharp Minimax Regret for Infinite-Memory Logistic Prediction ​
Author: Vaneet Aggarwal
Published: 8/28/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT
arXiv:2608.26515v1 Announce Type: cross Abstract: We study online prediction for a specific finite-alphabet, exogenously driven source with infinite input memory. Independent Rademacher inputs $(U_t)$ are observed sequentially, and the next binary mark has logit $\sum_{j=1}^{t}\theta_jU_{t+1-j}$, wh...
116. Physics-Informed Stochastic Configuration Machine: A Backpropagation-Free Neural Network with Fast Training for Nonlinear Differential Equations ​
Author: Yuehao Song (School of Automation, Central South University, Changsha, China), Zhong Chen (School of Automation, Central South University, Changsha, China), Lihui Cen (School of Automation, Central South University, Changsha, China), Liang Wu (Johns Hopkins University, Baltimore, USA), Kai Zhang (State Key Laboratory of Simulation and Regulation of Water Cycle in River Basin, China Institute of Water Resources and Hydropower Research, Beijing, China)
Published: 8/28/2026, 4:00:00 AM
Categories: math.NA, cs.AI, cs.LG, cs.NA, cs.SY, eess.SY
arXiv:2608.26549v1 Announce Type: cross Abstract: While Physics-Informed Neural Networks (PINNs) have emerged as a transformative paradigm for solving complex differential equations, their reliance on backpropagation-based gradient descent and automatic differentiation (AD) imposes significant compu...
117. Hadamard Flattening and Gaussian Pooling Sketch for Least Squares with Coordinate-wise Guarantee ​
Author: Zhao Song, Lichen Zhang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, stat.ML
arXiv:2608.26552v1 Announce Type: cross Abstract: Randomized sketch-and-solve algorithms accelerate overconstrained $\ell_2$ regression by replacing the input with a smaller problem. Standard subspace embeddings guarantee that the cost of the regression is nearly preserved, but coordinate-wise accur...
118. Dynamical phase selection controls compute scaling in looped transformers ​
Author: Gunn Kim
Published: 8/28/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cond-mat.stat-mech, cs.LG
arXiv:2608.26556v1 Announce Type: cross Abstract: A looped transformer performs inference by iterating a weight-tied map, making its computation a dynamical process whose cost is set by the resulting inference dynamics. Here we show that networks with identical architecture and objective, trained to...
119. hoBIT: A Profile-Aware Retrieval-Augmented Chatbot for University Academic Advising ​
Author: Yoonseo Kim, Seongmin Lee, Joongheon Kim, SeongKu Kang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.26604v1 Announce Type: cross Abstract: In university academic advising, identical questions can require different answers depending on a student's department, admission cohort, and degree program, causing profile-blind retrievers to surface plausible but inapplicable evidence. We present ...
120. A Unified Descriptive-Complexity Framework for Model Selection under Correlated Designs ​
Author: Yanhang Zhang, Wei Liu, Yuhong Yang
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.26618v1 Announce Type: cross Abstract: Model selection becomes particularly challenging under strong predictor dependence and model-class uncertainty, especially when there are exponentially many models. We propose a Descriptive-Complexity Information Criterion (DCIC) that regularizes lar...
121. Hierarchical Channel Stacking: A Structured Decision Framework for AI-Generated Image Detection ​
Author: Saifullah Shoaib, Akash Borigi, Rupendra Lekkala, Amaury Lendasse, Edward Ratner, Sai Sowjanya Bhamidipati, Alexander Schlager, Peggy Lindner
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.26648v1 Announce Type: cross Abstract: Many synthetic-image detectors produce accurate predictions but offer limited insight into how those decisions are formed. This paper introduces Hierarchical Channel Stacking (HCS), a compact framework for AI-generated image detection that converts i...
122. FOCUS & RePAIR: Mitigating Text Degeneration via Token-Level Guidance for Pruned Large Language Models ​
Author: Junyoung Lee, Sehyeon Park, Shinhyoung Jang, Seonha Ryu, Hojeong Kim, Hyunsei Lee, Il Hong Suh, Yeseong Kim
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.26676v1 Announce Type: cross Abstract: Pruning is a practical approach to compress large language models (LLMs), but it can amplify text degeneration, especially repetition loops, even when perplexity and task accuracy remain largely unchanged. In this work, we present a token-level analy...
123. SIGMA: Structured Noise-Effect-Aware Grouped Multi-Agent Aggregation ​
Author: Li Mingqian
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA
arXiv:2608.26683v1 Announce Type: cross Abstract: Cooperative multi-agent reinforcement learning (MARL) faces significant challenges in maintaining robust coordination under noisy observations. Although observation disturbances are often introduced independently across agents, their downstream effec...
124. Domain-Specific Self-Supervised Representation Learning for Retinal Fundus Classification ​
Author: Bekzat Nurlanbekova, Fung Fung Ting
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.26686v1 Announce Type: cross Abstract: Despite the growing number of public datasets, annotated medical images remain scarce. Supervised learning methods achieve strong performance on many benchmarks, however require large amounts of labeled data, which are costly and time-consuming to ob...
125. Daydreaming: Stealing Hidden Agent Skills through Black-Box Task Interaction ​
Author: Yu-Lin Tsai, Yu-An Lu, Ci-Yang Tsai, Muxi Lyu, Raluca Ada Popa, Chia-Mu Yu
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2608.26733v1 Announce Type: cross Abstract: Agent skills bundle instructions, reference data, and executable helpers that let a general agent perform specialized tasks. Hosted providers can keep these files secret while selling access to task results, making the skill itself a valuable target....
126. Generative Semantic Scene Completion ​
Author: Shi Chen, Weifeng Ge
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO
arXiv:2608.26737v1 Announce Type: cross Abstract: Outdoor LiDAR semantic scene completion (SSC) recovers a dense semantic voxel grid from a scan observing 1% of the target volume, under class imbalance beyond 7,000x. We recast SSC as generative semantic scene completion (GSSC): a single discrete-dif...
127. Equal Ranking Quality, Different Decisions: Training Order-Consistent LLM Scorers ​
Author: Markus Frohmann, Mahdiyar Alavi, Elizabeth Lingg, Navid Rekabsaz
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.IR, cs.LG
arXiv:2608.26762v1 Announce Type: cross Abstract: Rerankers, reward models and multi-document QA scorers score candidate documents or responses in one LLM prompt, so each score depends on their order. Such scorers are selected on ranking quality, but their scores determine a decision: what a score t...
128. Neural Renormalization Group Flow for Percolation ​
Author: Anaclara Alvez, Luca Camagna, Sergio Chibbaro, Cyril Furtlehner, Fran\c{c}ois Landes, Gianluca Manzan, Lorenzo Mensi
Published: 8/28/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cs.LG
arXiv:2608.26764v1 Announce Type: cross Abstract: Machine learning offers a possible route to data-driven real-space renormalization when the relevant observables are nonlocal and difficult to prescribe explicitly. We explore this idea for two-dimensional site percolation developping a supervised, s...
129. Incremental Recommendation via Causal Models ​
Author: Athanasios Vlontzos, David Gustafsson, Michael O'Riordan, Ciar'an M. Gilligan-Lee
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2608.26804v1 Announce Type: cross Abstract: Recommendation impressions are a finite resource, hence delivering a recommendation to a user who would discover the content organically yields no incremental value and displaces other recommendations that could. We address this by extending an exist...
130. Hyperspectral Diffusion Equivariant Imaging (HyDiff-EI): A Self-supervised Framework for Hyperspectral Image Inpainting ​
Author: Shuo Li, Mike Davies, Mehrdad Yaghoobi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.26812v1 Announce Type: cross Abstract: A novel Hyperspectral diffusion Equivariant Imaging (HyDiff-EI) framework for solving the hyperspectral image (HSI) inpainting problem has been presented here. Unlike conventional diffusion-based methods that rely on large-scale pretraining, HyDiff-E...
131. Bridging short- and medium-range weather forecasting with machine learning ​
Author: Timothy A. Smith, Mariah Pope, Sergey Frolov, Brett Basarab, Daniel Abdi, Paul Madden, Isidora Jankov
Published: 8/28/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG
arXiv:2608.26822v1 Announce Type: cross Abstract: The National Oceanic and Atmospheric Administration (NOAA) employs independent prediction systems for distinct forecast products. While some separation is practical, we argue that combining short- and medium-range weather into a single prediction sys...
132. Dose-PlanNet: Physics Based Radiotherapy Dose Prediction with Deep Learning ​
Author: Ankit Bhattacharjee, Sougata Maity, Santam Chakraborty, Indranil Mallick
Published: 8/28/2026, 4:00:00 AM
Categories: physics.med-ph, cs.CV, cs.LG
arXiv:2608.26901v1 Announce Type: cross Abstract: Automating prostate radiotherapy treatment planning is dosimetrically complex, particularly for extreme hypofractionated regimens. In this study, we introduce Dose-PlanNet, a physics-guided 3D deep learning architecture designed to predict dose distr...
133. Data-driven Koopman mode approximation: A neural power iteration algorithm ​
Author: Guillaume O. Berger, Rapha"el M. Jungers
Published: 8/28/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC
arXiv:2608.26943v1 Announce Type: cross Abstract: This paper proposes a novel data-driven algorithm to approximate the dominant eigenfunctions (aka.~modes) of the Koopman operator of nonlinear dynamical systems using neural networks. The relevance of learning the dominant Koopman modes is to approxi...
134. Squeezing More from Limited Data with Recursive Transformers ​
Author: Serdar G"ulbahar, Lukas Edman, Alexander Fraser
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.26973v1 Announce Type: cross Abstract: Pre-training under limited data requires a different view of scaling than web-scale language modeling. With a fixed data budget but relatively abundant compute, increasing parameter count helps only up to an optimal scale; beyond that point, models o...
135. Why not to use the Gaussian kernel ​
Author: Toni Karvonen, Chris J. Oates
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.NA, math.ST, stat.TH
arXiv:2608.26974v1 Announce Type: cross Abstract: Kernels measure similarity or correlation in tasks such as regression and classification. The Gaussian kernel, other names of which include squared exponential and radial basis function kernel, is one of the most popular in Gaussian process regressio...
136. Representation Measurements Under Function-Preserving Reparameterizations ​
Author: Abdullah Karasan
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.27020v1 Announce Type: cross Abstract: Hidden coordinates are not uniquely determined by a language model's input--output function, so representation-derived measurements should be invariant to function-preserving changes of basis. This study shows that column-permutation parallel analysi...
137. FoldPipe: Bounded Remote Streaming of Native Molecular Shards with Asynchronous Prefetch ​
Author: Dhiren Mukesh Khatri
Published: 8/28/2026, 4:00:00 AM
Categories: cs.PF, cs.LG
arXiv:2608.27029v1 Announce Type: cross Abstract: Training molecular machine-learning models on ephemeral or memory-constrained accelerator instances can require repeatedly retrieving preprocessed molecular graphs from remote storage. FoldPipe is a lightweight Python orchestration layer for already-...
138. Beyond Classification: Task-Dependent Learnability under Privacy-Motivated Image Transformations ​
Author: Leon Ranke, Wolfgang H"ubner, Ronny Hug, Michael Arens, J"urgen Beyerer
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.27066v1 Announce Type: cross Abstract: Privacy-Enhancing Technologies (PETs) in computer vision often rely on noise or image perturbations to protect visual data while securely processing it, creating a trade-off between task performance and protection. This trade-off is commonly evaluate...
139. Active Diffusion-Based Inference for Ill-Posed Inverse Problems under Incomplete Priors ​
Author: Jitao Xu, Nobuo Sato, Yaohang Li
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2608.27080v1 Announce Type: cross Abstract: Many scientific and engineering applications require estimating unknown parameters from experimentally observable data -- an inverse problem that is inherently challenging due to nonlinearity, noise, and ill-posedness. In this paper, we propose an ac...
140. SecureDrive-FL: Joint Differential Privacy and Gradient-Aware Selective Homomorphic Encryption for Federated Driver Monitoring ​
Author: Baran Can G"ul, Hanuma Siddhartha Tunuguntla, Anjana Arvind Naik, Abhishek Vijay Potekar, Nasser Jazdi, Michael Weyrich
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.27108v1 Announce Type: cross Abstract: Federated Learning (FL) enables privacy-aware distributed training, yet gradient updates remain exploitable: Man-in-the-Middle (MitM) interception exposes updates in transit, while model poisoning corrupts global convergence. We first introduce GASHE...
141. Linear Independence of Polynomial Compositions and Identifiability of Deep Neural Networks ​
Author: Kathl'en Kohn, Giovanni Luca Marchetti, Alex Massarenti, Massimiliano Mella
Published: 8/28/2026, 4:00:00 AM
Categories: math.AC, cs.LG, math.AG
arXiv:2608.27113v1 Announce Type: cross Abstract: Motivated by theoretical problems in deep learning, we conjecture that post-composing a fixed number of pairwise distinct nonconstant polynomials with a generic polynomial of sufficiently large degree yields linearly independent polynomials. This gen...
142. How AI Experiences Art: Emergent Aesthetic Structure in a Self-Supervised Multimodal Embedding Space ​
Author: Corey D. C. Heath
Published: 8/28/2026, 4:00:00 AM
Categories: cs.MM, cs.CV, cs.LG
arXiv:2608.27121v1 Announce Type: cross Abstract: Aesthetics are an important part of the symbolism of artistic works. Although subjective, humans categorize art based on the emotion evoked regardless of modality. What remains under-explored is how AI models form their own aesthetic categorization o...
143. Over-The-Air Extreme Learning Machines with Nonlinear Stacked Intelligent Metasurfaces ​
Author: Kyriakos Stylianopoulos, Mattia Fabiani, Giulia Torcolacci, Davide Dardari, George C. Alexandropoulos
Published: 8/28/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.27137v1 Announce Type: cross Abstract: The recently envisioned goal-oriented communications paradigm requires machine learning inference to be performed directly on wirelessly transferred data. This paper presents an eXtremely Large (XL) Multiple-Input Multiple-Output (MIMO) system that o...
144. Data-efficient crack quantification in lithium-ion cathodes using foundation model transfer ​
Author: Thorsten Tegetmeyer-Kleine, Thomas Schmitt, Phillip Aquino, Christiane Rahe, Dirk Uwe Sauer, Weihan Li
Published: 8/28/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.CV, cs.LG
arXiv:2608.27162v1 Announce Type: cross Abstract: Battery lifetime is central to sustainable electrification, yet the particle cracking that drives lithium-ion cathode aging is hard to measure: quantitative microscopy of this degradation is bottlenecked by annotation, because each destructive electr...
145. When Text Misleads: Inconsistent-Aware Reasoning for Audio-Grounded Dialogue ​
Author: Yen-Ju Lu, Yuzhe Wang, Yaohan Guan, Xiluo He, Jiarui Hai, Mingrui Liang, Kaavya Chaparala, Thomas Thebaud, Laureano Moro-Velazquez, Najim Dehak, Jesus Villalba
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, eess.AS
arXiv:2608.27176v1 Announce Type: cross Abstract: Understanding spoken dialogue requires joint reasoning over lexical content and paralinguistic acoustic signals such as emotion and conversational intent. However, existing evaluations often allow shortcuts based on transcripts or single-modality sol...
146. A Point-of-Prescription Safety-Check System for Adverse Drug Reactions in Rural Bangladeshi Hospitals: A Feasibility Study ​
Author: Shahir Abdullah
Published: 8/28/2026, 4:00:00 AM
Categories: cs.HC, cs.LG
arXiv:2608.27239v1 Announce Type: cross Abstract: Adverse drug reactions (ADRs) are a major, largely preventable source of patient harm. In high-income settings, electronic health records store a patient's allergy history and warn prescribers when a contraindicated drug is ordered; in rural Banglade...
147. Enforcing Dirichlet Boundary Conditions in Operator Learning ​
Author: Andrew M. Stuart, Margaret Trautner
Published: 8/28/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA
arXiv:2608.27256v1 Announce Type: cross Abstract: Operator learning in scientific machine learning is concerned with approximation of maps between infinite-dimensional function spaces; such maps frequently arise as the solution operators of partial differential equations (PDEs). Neural operators hav...
148. Recovering Expert Critic-Sourced Network Adjacency between Musical Artists from Acoustic Distributions: A Construct-Validity Approach ​
Author: Elena Badillo-Goicoechea, Fengfeng He
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.27291v1 Announce Type: cross Abstract: Music recommendation relies primarily on two signals: user-item interactions, which fail in the cold-start regime, and intrinsic musical content, available for any recording. We argue that a third, largely untapped signal is both richer and more prin...
149. LLMs Can Design Near-Optimal OR Algorithms ​
Author: Jackie Baek
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.27296v1 Announce Type: cross Abstract: We ask whether large language models (LLMs) can design effective algorithms for well-specified operations research (OR) problems. We study inventory control, queueing network control, and assortment optimization. We evaluate two levels of LLM use: at...
150. A Finite Sample Analysis for Quantile Temporal Difference Learning in Distributional Reinforcement Learning ​
Author: Zijie Cheng, Xiang Li, Yang Peng, Zhihua Zhang
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.27313v1 Announce Type: cross Abstract: We establish a global finite-sample guarantee for synchronous quantile temporal-difference learning (QTD) in tabular distributional reinforcement learning. The proof separates two stability mechanisms. A global comparison argument, based on the order...
151. Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090 ​
Author: Kairong Luo, Jiarui Cui, Yaorui Yin, Shengqi Chen, Yiming Yang, Linxiang Gao, Yanmohan Wang, Mingzhe Zhang, Kaiyue Wen, Kaifeng Lyu, Wenguang Chen
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.27370v1 Announce Type: cross Abstract: Language model pretraining has become almost synonymous with prohibitive cost, placing it out of reach for much of the academic and open-source communities. Although strong open-source efforts already exist, including open-weight models and open-sour...
152. Universality and sharp thresholds for ellipsoid fitting ​
Author: Frederic Koehler, Youngtak Sohn
Published: 8/28/2026, 4:00:00 AM
Categories: math.PR, cond-mat.dis-nn, cs.DS, cs.LG, math.ST, stat.TH
arXiv:2608.27372v1 Announce Type: cross Abstract: We establish a sharp phase transition for fitting random vectors by an ellipsoid. The random vectors have independent subgaussian coordinates with mean zero, variance one, and a common fourth moment, and the number of vectors is proportional to the s...
153. Token-Level Advertising ​
Author: Hanbing Liu, Bowei Zhang, Changyuan Yu, Yinyu Ye, Qi Qi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2608.27382v1 Announce Type: cross Abstract: Generative AI is transforming how people access information, challenging traditional advertising mechanisms built around predefined slots. Towards generation-native advertising, we propose the Latent Advertiser Mixture Auction (LAMA), a token-level a...
154. CorporateBench: Large-Scale Q&A Benchmarking with Temporal Knowledge Bases ​
Author: Sil Hamilton, Albert Yu Sun, Oscar J. Romero, Carl-Leander Henneking, David Mimno, Bishan Yang, Igor Labutov
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.IR, cs.LG
arXiv:2608.27391v1 Announce Type: cross Abstract: LLMs are increasingly able to answer complex questions about enterprise-scale document collections. But evaluation is hard: companies don't want to share internal communications, and synthetic datasets have been overly simple. We present CorporateBen...
155. How Language Models Organize and Structure Moral Knowledge ​
Author: Orion Reblitz-Richardson
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.27402v1 Announce Type: cross Abstract: How do large language models (LLMs) organize moral knowledge? Models detect moral content broadly, but detection is a low bar. We ask whether they go further, distinguishing moral foundations from one another and organizing the relationships between ...
156. Scaling Graph Neural Networks for Friend Recommendation: Multi-Hash User Embeddings and Temporal Neighbor Sampling ​
Author: Maksim Utushkin, Andrei Ovsiannikov, Alexander D'yakonov
Published: 8/28/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, cs.SI
arXiv:2608.27413v1 Announce Type: cross Abstract: Friend recommendation is inherently graph-structured: the relevance of a potential connection depends on multi-hop social context rather than user attributes alone. However, deploying message-passing GNNs on a production-scale social graph with hundr...
157. Learning a Continuous Sepsis Severity Score Without Hour-by-Hour Supervision: A Two-Site Retrospective Study ​
Author: Kevin Zhu, Ryan Zhang, Baraa Abed, Tilendra Choudhary, Malvern Madondo, Mehak Arora, Yixuan Yang, Alasdair Gent, Aditya Nagori, Omer T. Inan, Krista L. Haines, Patrick Georgoff, Suresh M. Agarwal, Vijay Krishnamoorthy, Tetsu Ohnuma, Mihai V. Podgoreanu, Michael R. Pinsky, Gilles Clermont, Craig M. Coopersmith, Craig S. Jabaley, Rishikesan Kamaleswaran
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.27421v1 Announce Type: cross Abstract: Currently used sepsis severity indices rely on fixed variables and weights established decades ago, which are coarsely discretized and calibrated to a cohort that no longer reflects contemporary critical care. No alternative learned directly from pat...
158. Federated Adversarial Training with Transformers ​
Author: Ahmed Aldahdooh, Wassim Hamidouche, Olivier D'eforges
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.CV
arXiv:2206.02131v2 Announce Type: replace Abstract: Federated learning (FL) has emerged to enable global model training over distributed clients' data while preserving its privacy. However, the global trained model is vulnerable to the evasion attacks especially, the adversarial examples (AEs), care...
159. Recurrent Reinforcement Learning with Memoroids ​
Author: Steven Morad, Chris Lu, Ryan Kortvelesy, Stephan Liwicki, Jakob Foerster, Amanda Prorok
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2402.09900v4 Announce Type: replace Abstract: Memory models such as Recurrent Neural Networks (RNNs) and Transformers address Partially Observable Markov Decision Processes (POMDPs) by mapping trajectories to latent Markov states. Neither model scales particularly well to long sequences, espec...
160. CollaFuse: Collaborative Diffusion Models ​
Author: Simeon Allmendinger, Domenique Zipperling, Lukas Struppek, Niklas K"uhl
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2406.14429v5 Announce Type: replace Abstract: In the landscape of generative artificial intelligence, diffusion-based models have emerged as a promising method for generating synthetic images. However, the application of diffusion models poses numerous challenges, particularly concerning data ...
161. Guided Data Generation for Understanding Model Behavior ​
Author: Eren Mehmet K{\i}ral, Nur\c{s}en Ayd{\i}n, \c{S}. .Ilker Birbil
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2502.06658v4 Announce Type: replace Abstract: We propose a method for generating distributions over the input space as an inspection tool for understanding trained models. Our framework poses questions of the form ``which inputs would make a trained model exhibit a specified behavior?'' and en...
162. MODIS: Multi-Omics Data Integration for Small and unpaired datasets ​
Author: Daniel Lepe-Soltero, Thierry Arti`eres, Ana"is Baudot, Paul Villoutreix
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2503.18856v3 Announce Type: replace Abstract: An important objective in computational biology is the efficient integration of multi-omics data. The task of integration comes with challenges: multi-omics data are most often unpaired (requiring diagonal integration), partially labeled with infor...
163. Plain Transformers Can be Powerful Graph Learners ​
Author: Liheng Ma, Soumyasundar Pal, Yingxue Zhang, Philip H. S. Torr, Mark Coates
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2504.12588v4 Announce Type: replace Abstract: Transformers have attained outstanding performance across various modalities, owing to their simple but powerful scaled-dot-product (SDP) attention mechanisms. Researchers have attempted to migrate Transformers to graph learning, but most advanced ...
164. From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning ​
Author: Yuzhen Huang, Weihao Zeng, Xingshan Zeng, Qi Zhu, Junxian He
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2505.22203v3 Announce Type: replace Abstract: Trustworthy verifiers are essential for the success of reinforcement learning with verifiable reward (RLVR), which is the core methodology behind various large reasoning models such as DeepSeek-R1. In complex domains like mathematical reasoning, ru...
165. Residual Reward Models: Leveraging Prior Knowledge for Efficient Preference-based Reinforcement Learning in Robotics ​
Author: Chenyang Cao, Miguel Rogel-Garc'ia, Mohamed Nabail, Xueqian Wang, Nicholas Rhinehart
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO
arXiv:2507.00611v2 Announce Type: replace Abstract: Preference-based Reinforcement Learning (PbRL) provides a promising alternative to heuristic reward design in complex robotic environments. However, PbRL often suffers from poor sample efficiency, requiring extensive and costly human feedback, whic...
166. AirLLM: Diffusion Policy-based Adaptive LoRA for Remote Fine-Tuning of LLM over the Air ​
Author: Shiyi Yang, Xiaoxue Yu, Rongpeng Li, Jianhang Zhu, Zhifeng Zhao, Honggang Zhang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2507.11515v2 Announce Type: replace Abstract: Operating Large Language Models (LLMs) on edge devices is increasingly challenged by limited communication bandwidth and strained computational and memory costs. Thus, cloud-assisted remote fine-tuning becomes indispensable. Nevertheless, existing ...
167. Provable one-poison backdoor attacks on linear models and ReLU neural networks ​
Author: Thorsten Peinemann, Paula Arnold, Sebastian Berndt, Thomas Eisenbarth, Esfandiar Mohammadi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2508.05600v3 Announce Type: replace Abstract: Backdoor poisoning attacks are a threat to machine learning models that are trained on data collected from untrusted sources; these attacks enable attackers to inject malicious behavior into the model that can be triggered by specially crafted inpu...
168. ReLATE: Accelerating Tensor Decomposition via Safe and Efficient Learning of Sparse Encodings ​
Author: Ahmed E. Helal, Fabio Checconi, Jan Laukemann, Yongseok Soh, Jesmin Jahan Tithi, Fabrizio Petrini, Jee Choi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.PF
arXiv:2509.00280v2 Announce Type: replace Abstract: Tensor decomposition (TD) is essential for analyzing high-dimensional sparse data, yet its irregular computations and memory-access patterns pose major performance challenges on modern parallel processors. Prior works rely on expert-designed sparse...
169. Leakage-Free Evaluation and Distribution-Robust Spatio-Temporal Graph Learning for Inductive Kriging ​
Author: Chen Yang, Changhao Zhao, Haoyang Zhao, Youquan He, Chen Wang, Jiansheng Fan
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.23631v2 Announce Type: replace Abstract: Inductive kriging estimates values at unobserved locations from sparse sensor data, enabling continuous field reconstruction when dense deployment is impractical. However, common 2 x 2 and 2 x 3 evaluation protocols can leak spatial information thr...
170. MCCE: A Framework for Multi-LLM Collaborative Search in Discrete Spaces with Similarity-Filtered Preference Learning ​
Author: Nian Ran, Zhongzheng Li, Yue Wang, Qingsong Ran, Xiaoyuan Zhang, Shikun Feng, Richard Allmendinger, Xiaoguang Zhao
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.06270v2 Announce Type: replace Abstract: Multi-objective discrete optimization problems, such as molecular design, pose significant challenges due to their vast and unstructured combinatorial spaces. Traditional evolutionary algorithms often get trapped in local optima, while expert knowl...
171. A Survey of LLM Prompt Datasets: Taxonomy, Linguistic Patterns, and Practical Uses ​
Author: Yuanming Zhang, Yan Lin, Arijit Khan, Huaiyu Wan
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2510.09316v3 Announce Type: replace Abstract: We compile 129 public LLM prompt datasets with more than 1.22TB and more than 673M instances and organize them into a unified taxonomy. We use seven datasets for detailed analysis and identify lexical, syntactic, and semantic patterns that distingu...
172. Absolute indices for determining compactness, separability and number of clusters ​
Author: Adil M. Bagirov, Ramiz M. Aliguliyev, Nargiz Sultanova, Sona Taheri
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2510.13065v3 Announce Type: replace Abstract: Finding "true" clusters in a data set is a challenging problem. Clustering solutions obtained using different models and algorithms do not necessarily provide compact and well-separated clusters or the optimal number of clusters. Cluster validity i...
173. The Principles of Diffusion Models ​
Author: Chieh-Hsin Lai, Yang Song, Dongjun Kim, Yuki Mitsufuji, Stefano Ermon
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GR
arXiv:2510.21890v3 Announce Type: replace Abstract: This book presents the core principles that have guided the development of diffusion models, tracing their origins and showing how diverse formulations arise from shared mathematical ideas. Diffusion modeling starts by defining a forward process th...
174. COFM: Consistent Optimal Transport Flow Matching via Partially Input Convex Neural Networks ​
Author: Fanghui Song, Zhongjian Wang, Jiebao Sun
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.06042v2 Announce Type: replace Abstract: Optimal transport (OT) provides a principled framework for learning mappings between probability distributions, and has found broad applications in generative modeling, inverse problems and scientific computing. Recently, flow matching methods have...
175. Diagnosing Conformal Prediction Failures Under Distribution Shift: A COVID-19 Case Study ​
Author: Chorok Lee
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2601.00908v2 Announce Type: replace Abstract: Conformal prediction provides distribution-free coverage guarantees, but these degrade under distribution shift - and practitioners lack tools to anticipate which deployed models will fail before observing test data. We propose SHapley Additive exP...
176. Private and interpretable clinical prediction with quantum-inspired tensor train models ​
Author: Jos'e Ram'on Pareja Monturiol, Juliette Sinnott, Roger G. Melko, Mohammad Kohandel
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, quant-ph
arXiv:2602.06110v2 Announce Type: replace Abstract: Publicly available clinical machine learning models pose an underappreciated privacy risk: their parameters or outputs can be exploited to recover information from patients whose data were used during training. Moreover, this risk is exacerbated by...
177. SynthCharge: An Electric Vehicle Routing Instance Generator with Feasibility Screening to Enable Learning-Based Optimization and Benchmarking ​
Author: Mertcan Daysalilar, Fuat Uyguroglu, Gabriel Nicolosi, Adam Meyers
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.03230v2 Announce Type: replace Abstract: The electric vehicle routing problem with time windows (EVRPTW) extends the classical VRPTW by introducing battery capacity constraints and charging station decisions. Existing benchmark datasets are often static and lack verifiable feasibility, wh...
178. Learning to Reason with Curriculum I: Provable Benefits of Autocurriculum ​
Author: Nived Rajaraman, Audrey Huang, Miro Dudik, Robert Schapire, Dylan J. Foster, Akshay Krishnamurthy
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2603.18325v2 Announce Type: replace Abstract: Chain-of-thought reasoning, where language models expend additional computation by producing thinking tokens prior to final responses, has driven significant advances in model capabilities. However, training these reasoning models is extremely cost...
179. Stable but Wrong: When Learning Stabilizes Away from the Truth ​
Author: Zhipeng Zhang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.21491v2 Announce Type: replace Abstract: Stable training is often treated as evidence that learning is succeeding, but stability characterizes optimization behavior rather than correctness relative to an external objective. We study what happens when the signal being optimized remains per...
180. Active Preference Learning over Latent Preference Archetypes for Many-Objective Bayesian Optimization ​
Author: Manisha Dubey, Sebastiaan De Peuter, Wanrong Wang, Samuel Kaski
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2603.28410v2 Announce Type: replace Abstract: Preference-based many-objective Bayesian optimization typically assumes that all pairwise comparisons arise from a single latent utility function, despite real decision makers often exhibiting multiple latent preference archetypes across contexts. ...
181. The Rashomon Effect for Visualizing High-Dimensional Data ​
Author: Yiyang Sun, Haiyang Huang, Gaurav Rajesh Parikh, Cynthia Rudin
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.00485v3 Announce Type: replace Abstract: Dimension reduction (DR) is inherently non-unique: multiple embeddings can preserve the structure of high-dimensional data equally well while differing in layout or geometry. In this paper, we formally define the Rashomon set for DR -- the collecti...
182. $p1$: Better Prompt Optimization with Fewer Prompts ​
Author: Zhaolin Gao (Sid), Yu (Sid), Wang, Bo Liu, Thorsten Joachims, Kiant'e Brantley, Wen Sun
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2604.08801v2 Announce Type: replace Abstract: Prompt optimization improves language models without updating their weights by searching for a better system prompt, but its effectiveness varies widely across tasks. We study what makes a task amenable to prompt optimization. We show that the rewa...
183. Can LLMs Accurately Score Medical Diagnoses and Clinical Reasoning? ​
Author: Amy Rouillard, Sitwala Mundia, Linda Camara, Ziyaad Dangor, Michael Cameron Gramanie, Ismail Kalla, Shabir A. Madhi, Kajal Morar, Marlvin T. Ncube, Haroon Saloojee, Bruce A. Bassett
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2604.14892v4 Announce Type: replace Abstract: Evaluating medical AI systems using expert clinician panels is costly and slow, motivating the use of large language models (LLMs) as alternative adjudicators. Here, we evaluate an LLM Jury, composed of three frontier AI models, for scoring 3334 di...
184. A unified convergence theory for adaptive first-order methods in the nonconvex case, including AdaNorm, full and diagonal AdaGrad and Muon ​
Author: S. Gratton, Ph. L. Toint
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.17423v3 Announce Type: replace Abstract: A unified framework for first-order optimization algorithms fornonconvex unconstrained optimization is proposed that uses adaptivelypreconditioned gradients and includes popular methods such as full anddiagonal AdaGrad, AdaNorm, as well as an adpat...
185. Aitchison Embeddings for Learning Compositional Graph Representations ​
Author: Nikolaos Nakis, Chrysoula Kosma, Panagiotis Promponas, Michail Chatzianastasis, Giannis Nikolentzos
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2605.00716v3 Announce Type: replace Abstract: Representation learning is central to graph machine learning, powering tasks such as link prediction and node classification. However, most graph embeddings are hard to interpret, offering limited insight into how learned features relate to graph s...
186. Cartan flow matching ​
Author: Francesco Ruscelli, Ferdinando Zanchetta, Rita Fioresi
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.03588v2 Announce Type: replace Abstract: We introduce Cartan flow matching, a general framework for training flow matching models on Riemannian symmetric spaces, i.e. Riemannian manifolds with the property that at any point there exists a geodesic symmetry. This is a large class of manifo...
187. HINT-SD: Targeted Hindsight Self-Distillation for Long-Horizon Agents ​
Author: Woongyeong Yeo, Yumin Choi, Taekyung Ki, Sung Ju Hwang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2605.17873v2 Announce Type: replace Abstract: Training long-horizon LLM agents with reinforcement learning is challenging because sparse outcome rewards reveal whether a task succeeds, but not which intermediate actions caused the outcome or how they should be corrected. Recent methods allevia...
188. GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets ​
Author: Zhangyang Yao, Haiyan Zhao, Haoyu Wang, Xu Han
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.18475v2 Announce Type: replace Abstract: Mixed-precision quantization improves the budget--accuracy trade-off for large language models (LLMs) by allocating more bits to sensitive modules. However, automating this allocation at LLM scale faces a unique combination of constraints: learnabl...
189. The Attribution Contract for Generative Language Models ​
Author: Giang Nguyen
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.23080v3 Announce Type: replace Abstract: Feature attribution scores each part of an input by how much it explains a model's output. We argue that in generative language models these scores carry no fixed meaning. A classifier has a single output to explain, but a generative model produces...
190. LoopMoE: Unifying Iterative Computation with Mixture-of-Experts for Language Modeling ​
Author: Wenkai Chen, Tianshu Li, Wenyong Huang, Yichun Yin, Lifeng Shang, Chengwei Qin
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.04438v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) and looped architectures scale models along two orthogonal axes, namely parameter capacity and effective depth. However, mainstream looped architectures rely on dense backbones that couple parameter count with per-token FLO...
191. Beyond FLOPs: Benchmarking Real Inference Acceleration of LLM Pruning under a GEMM-Centric Taxonomy ​
Author: Haozhe Hu, Hao Wu, Anhao Zhao, Longwei Ding, Peiran Yin, Yunpu Ma, Xiaoyu Shen
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2606.09080v2 Announce Type: replace Abstract: Pruning has emerged as a dominant paradigm for accelerating large language model (LLM) inference, spanning a broad spectrum of methods that remove computation across tokens, layers, heads, dimensions, and attention patterns. Despite sharing the sam...
192. When Do Autoregressive Sequence Models Forecast Physical Wavefields? A Controlled Study on Synthetic Seismograms ​
Author: Waleed Esmail, Stuart Russell, Jana Klinge, Alexander Kappes, Christine Thomas
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.IM
arXiv:2606.10868v2 Announce Type: replace Abstract: Long-horizon autoregressive forecasting of oscillatory physical signals, such as seismograms, gravitational-wave strain, and similar wavefields is limited by error accumulation: as a causal model is fed its own outputs over hundreds of steps, small...
193. You Don't Need to Run Every Eval ​
Author: Yuchen Zeng, Dimitris Papailiopoulos
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.24020v2 Announce Type: replace Abstract: A modern model release reports scores on 40+ benchmarks and the same evaluations were run many more times before it: to track training progress, compare design choices, and select the checkpoint for the release. But do we need to run every eval? We...
194. Tensorion: A Tensor-Aware Generalization of the Muon Optimizer ​
Author: Vladimir Bogachev, Vladimir Aletov, Alexander Molozhavenko, Sergei Kudriashov, Maxim Rakhuba
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.NA, math.NA
arXiv:2606.25975v2 Announce Type: replace Abstract: Common first-order optimizers, such as Adam, implicitly treat each parameter block as an unstructured vector, which disregards the multilinear weight structure present in many modern machine learning models. Recent work has shown that exploiting ma...
195. Curating Same-Family Neural Networks for LLM-Guided Model Improvement: A Controlled Case Study ​
Author: Kabir Dev Paul Baghel, Radu Timofte, Dmitry Ignatov
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.05704v2 Announce Type: replace Abstract: Neural-network repositories contain executable models, recipes, input transformations, and measured accuracies. We study whether one same-family experiment can be curated as prompt guidance for LLM-based improvement of a low-performing target under...
196. Auditing Invisible Weight Updates with Reference Traces ​
Author: Zekai Shang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.09800v4 Announce Type: replace Abstract: Direct low-precision write-back can erase nonzero optimizer proposals. We ask what a high-precision reference trace establishes before a low-precision run. The exact target-code event is auditable coordinatewise on a realized target trajectory; pre...
197. Neural Non-Equilibrium Hamiltonian Monte Carlo for Corrected Boltzmann Sampling ​
Author: Moxian Qian
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, hep-lat
arXiv:2607.15682v2 Announce Type: replace Abstract: Sampling from an unnormalized Boltzmann density requires proposals that move probability mass globally while retaining enough path-probability information for statistical correction. We introduce Neural Non-Equilibrium Hamiltonian Monte Carlo (NHMC...
198. Drift-Adaptive ICU Intervention Prediction: Freezing the Physiological Encoder for Auditable Model Updating ​
Author: Fatema Ferdous Tamanna, K. M. Merajul Arefin, Md. Abdul Masud
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR, q-bio.QM
arXiv:2607.19020v3 Announce Type: replace Abstract: Clinical decision support degrades as treatment protocols evolve, but the obstacle to updating a deployed model is governance as much as accuracy: once retraining touches every parameter, no one can say afterwards where the update acted. We propose...
199. SAGA: Score-Weighted Adaptive Generation Alignment for Low-Resource Nordic Language Models ​
Author: Hoda Fakharzadehjahromy, Emil Wiman, Andreas Bueff, Hafsteinn Einarsson, Fredrik Heintz
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.06179v2 Announce Type: replace Abstract: Preference optimisation has proven effective for improving large language models but typically relies on costly human preference annotations. Extending these methods to morphologically rich, low-resource languages remains challenging because such a...
200. REOPD: Reliability-Adaptive Reward Extrapolation for On-Policy Distillation ​
Author: Yang Sun, Lichao Ma, Houyuan Qin, Yuxin Liu, Hanyang Lu, Yao Zhu, Pinlong Cai, Guohang Yan
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11698v3 Announce Type: replace Abstract: On-policy distillation (OPD) trains a student on its own trajectories under dense token-level supervision from a teacher. Reward-extrapolation methods such as ExOPD amplify the teacher-reference log-likelihood ratio to move beyond direct imitation,...
201. NRCD: An Open Database of Collegiate Running with Unified Performance Standardization ​
Author: Jonathan A. Karr Jr., Ryan M. Fryer, Ben Darden, Nicholas Pell, Kayla Ambrose, Evan Hall, Ramzi K. Bualuan, Nitesh V. Chawla
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CY, cs.IR
arXiv:2608.14776v2 Announce Type: replace Abstract: Collegiate running in the United States generates thousands of race results annually in cross country and track and field, yet no large-scale dataset has been publicly available for research. Existing websites such as Athletic.net, MileSplit, and T...
202. A Real-Time Tsetlin Machine-based Non-intrusive Load Monitoring System on MCUs ​
Author: Han Wu, Tianhang Tan, Shengyu Duan, Alex Yakovlev, Rishad Shafik, Tousif Rahman
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.18780v2 Announce Type: replace Abstract: Non-Intrusive Load Monitoring (NILM) systems estimate individual appliance energy consumption from a single aggregate meter, without requiring separate sensors for each device. By installing a single meter that measures a building's total electrici...
203. Rethinking Expressivity and Efficiency in Test-Time Training ​
Author: Zeyun Zhong, Joya Chen, Manuel Martin, Frederik Diederichs, Juergen Gall, Juergen Beyerer
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.21308v2 Announce Type: replace Abstract: Test-Time Training (TTT) enables long-context processing via continuous weight updates during inference, but current methods struggle to balance the expressivity of per-token update dynamics with the hardware efficiency of chunk-wise approximations...
204. Decoupled Physical Modeling and Execution for Physics Reasoning ​
Author: Ye Zhang, Xuehang Guo, Rui Pan, Pengfei Yu, Denghui Zhang, Manling Li, Qingyun Wang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.22126v2 Announce Type: replace Abstract: Physics reasoning requires constructing a consistent model of the underlying physical system rather than relying solely on symbolic or formula-based manipulation. Although large language models have shown strong ability in solving math and coding p...
205. Learning Generalizable Behaviors for Terminal Agents ​
Author: Yihang Yao, Bo Pang, Xuan Phi Nguyen, Ding Zhao, Shafiq Joty, Semih Yavuz
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.22631v2 Announce Type: replace Abstract: Terminal agents are a compelling application of large language models (LLMs), with the potential to integrate deeply into users' daily workflows. Reinforcement learning (RL) is a key technique for improving their capabilities, making scalable train...
206. Mol-JEPA: A multimodal Joint Embedding Predictive Architecture for Molecules ​
Author: Florian Rottach, Sebastian Schieferdecker, William Rudman, Randall Balestriero, Carsten Eickhoff
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.22642v2 Announce Type: replace Abstract: Despite recent advances in molecular foundation models, several limitations remain, such as chemically invalid augmentations, modality collapse, and incomplete representation of biochemical environments. To address these challenges, we present \tex...
207. CAT-GS: Balanced Multimodal Learning via Calibrated Gating and Fusion Surgery ​
Author: Mahir Shahriar Tamim, Sharjil Khan, Md. Samiul Alim, Tanvir Ahmed Khan, Shafin Rahman, Nabeel Mohammed
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24947v2 Announce Type: replace Abstract: End-to-end training of multimodal neural networks often exhibits unstable neural dynamics characterized by three coupled failure modes that degrade learning: (i) modality imbalance, where one branch dominates gradient-based optimization; (ii) unsta...
208. A Token-Level Analysis of Sampled-Token Reverse-KL On-Policy Distillation ​
Author: Bing Shao, Jiazheng Zhang, Long Ma, Yujiong Shen, Senjie Jin, Xin Guo, Yuming Yang, Mingxu Chai, Zhiheng Xi, Boyang Liu, Junlin Shang, Tao Gui, Qi Zhang, Xuanjing Huang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.25643v2 Announce Type: replace Abstract: On-policy distillation (OPD) supervises a student on its own trajectories with token-level signals from a frozen teacher, yet how a sampled loss allocates updates across tokens remains poorly understood. We analyze the gradient of the per-token K2 ...
209. Group-Shared Low-Rank Approximation for Mobile-Efficient Pointwise Convolutions in Large-Kernel CNNs ​
Author: Hao Luo, Yiting Yang, Wenyi Zhao, Man Jiang, Zhijun Lin, Ghulam Mohiuddin, Ting Jiang, Kunming Luo, Zihao Zhang, Qingsen Yan, Guoqing Wang, Wei Dong, Peng Wang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.26069v2 Announce Type: replace Abstract: Large-kernel Convolutional Neural Networks (CNNs) deliver remarkable performance in vision tasks by significantly expanding receptive fields, yet their quadratic parameter growth critically impedes storage-efficient edge deployment. While existing ...
210. Bregman Linearized Augmented Lagrangian Method for Nonconvex Constrained Stochastic Zeroth-order Optimization ​
Author: Qiankun Shi, Han Yuan, Xiao Wang, Hao Wang
Published: 8/28/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2504.09409v2 Announce Type: replace-cross Abstract: In this paper, we study nonconvex constrained stochastic zeroth-order optimization problems, for which we have access to exact information of constraints and noisy function values of the objective. We propose a Bregman linearized augmented La...
211. STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation ​
Author: Hossein Goli, Michael Gimelfarb, Nathan Samuel de Lara, Haruki Nishimura, Masha Itkina, Florian Shkurti
Published: 8/28/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2505.20781v2 Announce Type: replace-cross Abstract: Off-policy evaluation (OPE) estimates the performance of a target policy using offline data collected from a behavior policy, and is crucial in domains such as robotics or healthcare where direct interaction with the environment is costly or ...
212. Gaussian Processes and Reproducing Kernel Hilbert Spaces: Connections and Equivalences ​
Author: Motonobu Kanagawa, Philipp Hennig, Dino Sejdinovic, Bharath K. Sriperumbudur
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.NA, math.PR, math.ST, stat.TH
arXiv:2506.17366v2 Announce Type: replace-cross Abstract: This monograph studies the relations between two approaches using positive definite kernels: probabilistic methods using Gaussian processes, and non-probabilistic methods using reproducing kernel Hilbert spaces (RKHS). They are widely studied...
213. Refine-POI: Reinforcement Fine-Tuned Large Language Models for Next Point-of-Interest Recommendation ​
Author: Peibo Li, Shuang Ao, Hao Xue, Yang Song, Maarten de Rijke, Johan Barth'elemy, Tomasz Bednarz, Flora D. Salim
Published: 8/28/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG
arXiv:2506.21599v5 Announce Type: replace-cross Abstract: Advancing large language models (LLMs) for the next point-of-interest (POI) recommendation task faces two fundamental challenges: (i) although existing methods produce semantic IDs that incorporate semantic information, their topology-blind i...
214. Pushing the Envelope of LLM Inference with Ultra-Low-Bit Quantized Models ​
Author: Evangelos Georganas, Dhiraj Kalamkar, Alexander Heinecke, Pradeep Dubey
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.PF
arXiv:2508.06753v3 Announce Type: replace-cross Abstract: The advent of ultra-low-bit LLM models, approaching the perplexity and task accuracy of their full precision counterparts, is ushering in a new era of LLM inference. While these advances promise models that are cost-effective regarding latenc...
215. Stack Trace-Based Crash Deduplication with Transformer Adaptation ​
Author: Md Afif Al Mamun, Gias Uddin, Lan Xia, Longyu Zhang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.SE, cs.LG
arXiv:2508.19449v2 Announce Type: replace-cross Abstract: Automated crash reporting systems generate large volumes of duplicate reports, overwhelming issue-tracking systems and increasing developer workload. Traditional stack trace-based deduplication methods---relying on string similarity, rule-bas...
216. Sequential Additivity in Distributionally Robust Ranking and Selection ​
Author: Zaile Li, Yuchen Wan, L. Jeff Hong
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2509.06147v2 Announce Type: replace-cross Abstract: Ranking and selection (R&S) seeks to identify the alternative with the best mean performance from a finite collection of simulated alternatives. Its practical value depends on accurate simulation input modeling, which is often hindered by inp...
217. LLM Analysis of 150+ years of German Parliamentary Debates on Migration Reveals Shift from Post-War Solidarity to Anti-Solidarity in the Last Decade ​
Author: Aida Kostikova, Ole P"utz, Steffen Eger, Olga Sabelfeld, Benjamin Paassen
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.CY, cs.LG
arXiv:2509.07274v4 Announce Type: replace-cross Abstract: Migration has been a core topic in German political debate, from postwar expellee displacement to labor migration and recent refugee movements. Large-scale analysis of such political discourse has traditionally required extensive manual annot...
218. A Framework for Low-Effort Training Data Generation for Urban Semantic Segmentation ​
Author: Damjan Kal\v{s}an, Denis Zavadski, Tim K"uchler, Haebom Lee, Stefan Roth, Carsten Rother
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.LG
arXiv:2510.11567v2 Announce Type: replace-cross Abstract: Synthetic datasets are widely used for training urban scene recognition models, but even highly realistic renderings show a noticeable gap to real imagery. This gap is particularly pronounced when adapting to a specific target domain, such as...
219. UCB for Large-Scale Pure Exploration: Beyond Sub-Gaussianity ​
Author: Zaile Li, Weiwei Fan, L. Jeff Hong
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2511.22273v2 Announce Type: replace-cross Abstract: Selecting the best alternative from a finite set is the central objective of ranking and selection (R&S) and best arm identification (BAI). Traditional R&S or BAI approaches have predominantly relied on Gaussian or sub-Gaussian assumptions on...
220. Quantitative mapping from conventional MRI using self-supervised physics-guided deep learning: applications to a large-scale, clinically heterogeneous dataset ​
Author: Jelmer van Lune, Stefano Mandija, Oscar van der Heide, Matteo Maspero, Martin B. Schilder, Jan Willem Dankbaar, Cornelis A. T. van den Berg, Alessandro Sbrizzi
Published: 8/28/2026, 4:00:00 AM
Categories: physics.med-ph, cs.CV, cs.LG
arXiv:2601.05063v2 Announce Type: replace-cross Abstract: Magnetic resonance imaging (MRI) is a cornerstone of clinical neuroimaging, yet conventional MRIs provide qualitative information heavily dependent on scanner hardware and acquisition settings. While quantitative MRI (qMRI) offers intrinsic t...
221. A Flexible Empirical Bayes Approach to Generalized Linear Models, with Applications to Sparse Logistic Regression ​
Author: Dongyue Xie, Matthew Stephens
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO, stat.ME
arXiv:2601.21217v2 Announce Type: replace-cross Abstract: We introduce a flexible empirical Bayes approach for fitting Bayesian generalized linear models. Specifically, we adopt a novel mean-field variational inference (VI) method and the prior is estimated within the VI algorithm, making the method...
222. Toward all-optical unsupervised Hebbian learning in deep photonic neuromorphic networks ​
Author: Xi Li, Disha Biswas, Peng Zhou, Wesley H. Brigner, Anna Capuano, Joseph S. Friedman, Qing Gu
Published: 8/28/2026, 4:00:00 AM
Categories: physics.optics, cond-mat.dis-nn, cs.ET, cs.LG
arXiv:2601.22300v4 Announce Type: replace-cross Abstract: We propose a deep photonic neuromorphic network (PNN) architecture based on phase-change material (PCM) synapses and local optical feedback for online, unsupervised Hebbian learning. The proposed architecture combines optical vector-matrix mu...
223. Modular Expert Merging for Biomedical Retrieval ​
Author: Sameh Khattab, Jean-Philippe Corbeil, Osman Alperen \c{C}inar-Kora\c{s}, Amin Dada, Julian Friedrich, Jiawei He, Douglas Teodoro, Jens Kleesiek
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2602.04731v2 Announce Type: replace-cross Abstract: Adapting general-purpose LLMs into domain-specialized dense retrievers typically requires large-scale training on mixed-domain data. We show that merging independently trained domain-specialized experts consistently exceeds this approach acro...
224. A Very Big Video Reasoning Suite ​
Author: Maijunxian Wang, Ruisi Wang, Juyi Lin, Ran Ji, Thadd"aus Wiedemer, Qingying Gao, Dezhi Luo, Yaoyao Qian, Lianyu Huang, Zelong Hong, Jiahui Ge, Qianli Ma, Hang He, Yifan Zhou, Lingzi Guo, Lantao Mei, Jiachen Li, Hanwen Xing, Tianqi Zhao, Fengyuan Yu, Weihang Xiao, Yizheng Jiao, Jianheng Hou, Danyang Zhang, Pengcheng Xu, Boyang Zhong, Zehong Zhao, Gaoyun Fang, John Kitaoka, Yile Xu, Hua Xu, Kenton Blacutt, Tin Nguyen, Siyuan Song, Haoran Sun, Shaoyue Wen, Linyang He, Runming Wang, Yanzhi Wang, Mengyue Yang, Ziqiao Ma, Rapha"el Milli`ere, Freda Shi, Nuno Vasconcelos, Daniel Khashabi, Alan Yuille, Yilun Du, Ziming Liu, Bo Li, Dahua Lin, Ziwei Liu, Vikash Kumar, Yijiang Li, Lei Yang, Zhongang Cai, Hokin Deng
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MM, cs.RO
arXiv:2602.20159v3 Announce Type: replace-cross Abstract: Rapid progress in video models has largely focused on visual quality, leaving their reasoning capabilities underexplored. Video reasoning grounds intelligence in spatiotemporally consistent visual environments that go beyond what text can nat...
225. PACIFIER: Pacing Opinion Depolarization via a Unified Graph Learning Framework ​
Author: Mingkai Liao, Zijun Zhang, Mingyang Zhou
Published: 8/28/2026, 4:00:00 AM
Categories: cs.SI, cs.LG
arXiv:2602.23390v4 Announce Type: replace-cross Abstract: Online social networks often form opposing echo chambers that reinforce opinion polarization. Under the Friedkin-Johnsen (FJ) model, ModerateInternal (MI) and ModerateExpressed (ME) reduce polarization by neutralizing selected users' internal...
226. Bayes with No Shame: Admissibility Geometries of Predictive Inference ​
Author: Nicholas G. Polson, Daniel Zantedeschi
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2603.05335v4 Announce Type: replace-cross Abstract: Modern predictive systems combine predictors, sequential monitors, prediction sets, and online strategies, each with a different certificate of optimality. We study four criterion-relative geometries: Blackwell risk dominance, anytime-valid a...
227. Global universality via discrete-time signatures ​
Author: Mihriban Ceylan, David J. Pr"omel
Published: 8/28/2026, 4:00:00 AM
Categories: math.PR, cs.LG, q-fin.MF
arXiv:2603.09773v2 Announce Type: replace-cross Abstract: We establish global universal approximation theorems for non-anticipative and general path-dependent functionals on spaces of piecewise linear paths, stating that linear functionals of the corresponding signatures are dense with respect to $L...
228. Learning to Predict, Discover, and Reason in High-Dimensional Event Sequences ​
Author: Hugo Math
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2603.16313v3 Announce Type: replace-cross Abstract: Electronic control units (ECUs) embedded within modern vehicles generate a large number of asynchronous events known as diagnostic trouble codes (DTCs). These discrete events form complex temporal sequences that reflect the evolving health of...
229. Grounded Token Initialization for New Vocabulary in LMs for Generative Recommendation ​
Author: Daiwei Chen, Zhoutong Fu, Chengming Jiang, Haichao Zhang, Ran Zhou, Tan Wang, Chunnan Yao, Guoyao Li, Rui Cai, Yihan Cao, Ruijie Jiang, Fedor Borisyuk, Jianqiang Shen, Jingwei Wu, Ramya Korlakai Vinayak
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2604.02324v2 Announce Type: replace-cross Abstract: Language models (LMs) are increasingly extended with new learnable vocabulary tokens for domain-specific tasks, such as Semantic-ID tokens in generative recommendation. The standard practice initializes these new tokens as the mean of existin...
230. MOMO: A framework for seamless physical, verbal, and graphical robot skill learning and adaptation ​
Author: Markus Knauer, Edoardo Fiorini, Maximilian M"uhlbauer, Stefan Schneyer, Promwat Angsuratanawech, Florian Samuel Lay, Timo Bachmann, Samuel Bustamante, Korbinian Nottensteiner, Freek Stulp, Alin Albu-Sch"affer, Jo~ao Silv'erio, Thomas Eiband
Published: 8/28/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CL, cs.HC, cs.LG
arXiv:2604.20468v3 Announce Type: replace-cross Abstract: Industrial robot applications require increasingly flexible systems that non-expert users can easily adapt for varying tasks and environments. However, different adaptations benefit from different interaction modalities. We present an interac...
231. MambaCSP: Hybrid-Attention State Space Models for Hardware-Efficient Channel State Prediction ​
Author: Aladin Djuhera, Haris Gacanin, Holger Boche
Published: 8/28/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, eess.SP, math.IT
arXiv:2604.21957v2 Announce Type: replace-cross Abstract: Recent works have demonstrated that attention-based transformer and large language model (LLM) architectures can achieve strong channel state prediction (CSP) performance by capturing long-range temporal dependencies across channel state info...
232. Leveraging Code Automorphisms for Improved Syndrome-Based Neural Decoding ​
Author: Rapha"el Le Bidan, Ahmad Ismail, Elsa Dupraz, Charbel Abdel Nour
Published: 8/28/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT
arXiv:2605.03620v2 Announce Type: replace-cross Abstract: Syndrome-based neural decoding (SBND) has emerged as a promising deep learning approach for soft-decision decoding of high-rate, short-length codes. However, this approach still has substantial room for improvement. In this paper, we show how...
233. From Accuracy to Auditability: A Survey of Determinism in Financial AI Systems ​
Author: Ruizhe Zhou, Xiaoyang Liu, Gaoyuan Du, Yi Zheng, Shouxi Ren, Deepayan Chakrabarti, Dengdu Jiang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.DC, cs.LG, cs.SI, q-fin.CP
arXiv:2605.23955v4 Announce Type: replace-cross Abstract: Deploying machine learning in regulated financial environments -- credit risk, fraud detection, and anti-money laundering -- exposes critical vulnerabilities in algorithmic reproducibility. While early financial ML addressed statistical chall...
234. Out of Sight, Not Out of Mind: Unveiling Latent Attack in Latent-based Multi-Agent Systems ​
Author: Chenxi Wang, Ruiyang Huang, Jiayan Sun, Lei Wei, Yifan Wu
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.MA
arXiv:2605.28214v2 Announce Type: replace-cross Abstract: Latent-based multi-agent systems replace parts of explicit inter-agent communication with hidden representations, offering a new direction for efficient and flexible agent collaboration. However, moving coordination into latent space may also...
235. Adaptive Inference for Resource-Constrained Dynamic Pricing ​
Author: Ruicheng Ao, Jiashuo Jiang, David Simchi-Levi
Published: 8/28/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2606.03736v3 Announce Type: replace-cross Abstract: We study dynamic pricing over a finite selling horizon when limited resource capacity determines revenue and the observations available for inference at a prespecified price. Resource depletion can remove the target neighborhood from the feas...
236. Clinically Aligned Geometry Constraints for Robust IVUS Vessel Boundary Segmentation ​
Author: Yunshu Chen, Litao Yang, Giuseppe Di Giovanni, Jordan Tan, Deval Mehta, Andrew Lin, Derek Chew, Masasi Fujino, Julie Butters, Stephen Nicholls, Zongyuan Ge, Kyung Hoon Cho
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2606.18723v2 Announce Type: replace-cross Abstract: Intravascular ultrasound (IVUS) lumen and external elastic membrane (EEM) segmentation is important for quantitative coronary plaque burden assessment. Errors in lumen or EEM delineation directly propagate to plaque area, plaque burden and ge...
237. When Top-K Misses the Decision: Tool-Call Drift in Multi-Teacher On-Policy Distillation ​
Author: Jiabin Shen, Guang Chen, Chengjun Mao
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.07050v5 Announce Type: replace-cross Abstract: Top-$K$ teacher logits make on-policy distillation tractable, but retained teacher mass does not certify student-relative gradient fidelity. We study routed, two-teacher tool-use distillation. In a frozen Qwen3.5-9B audit, the tool teacher ra...
238. Gradient-free learning of a closed-loop wall controller for turbulent drag reduction ​
Author: Giorgio Maria Cavallazzi, Miguel P'erez Cuadrado, Alfredo Pinelli
Published: 8/28/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG
arXiv:2607.12626v2 Announce Type: replace-cross Abstract: Closed-loop wall controllers learnt by multi-agent reinforcement learning are usually trained on periodic boxes far smaller than the flows they are meant to drive, and a large part of their drag reduction is lost when they are carried across....
239. ATLAS: Automated Approximation of Transformers for Efficient Homomorphic Inference in One Hour ​
Author: Jianhang Xie, Sicheng Tan, Vishnu Naresh Boddeti, Zhichao Lu
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.23478v2 Announce Type: replace-cross Abstract: Fully homomorphic encryption (FHE) lets a server run inference on encrypted data with strong privacy guarantees, but running a Transformer under FHE is expensive. Its non-linear operations, such as softmax, normalization, and activation, must...
240. TriShieldRAG: 3 Rings, One Blind Spot in Layered Defenses for Retrieval-Augmented Generation ​
Author: Susil Kumar Mohanty, Rohit Patel, Kosuru Yuvaraj, Jeenal Chaudhary, Disha Singhania
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG
arXiv:2607.23838v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) grounds LLM answers in query-time retrieved documents, so reliability depends on what the retriever returns. PoisonedRAG (Zou et al., USENIX Security'25) showed five crafted documents mislead an undefended...
241. When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs ​
Author: Jiaming Cheng, Subhransu Das, Rajiv Ramnath
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2608.04893v2 Announce Type: replace-cross Abstract: Multi-agent LLM systems relay key-value caches instead of text and credit their gains to exchanged "latent thoughts". That credit is a claim about which example's cache is relayed, not merely that one is. We audit it causally in released syst...
242. MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification ​
Author: Daniel Perkins, John Squires, Janou Milligan, Chandra Raskoti, Linda Ungerboeck
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG
arXiv:2608.13463v2 Announce Type: replace-cross Abstract: Modern image classification models excel when trained on single task-specific datasets but often struggle to generalize across domains and difficulty levels. We propose ARMDIL, an Adaptive Router for Multi-Domain Image Classification with LLM...
243. VQC-ZTI: Variational Quantum Control for Zero Trust Protection of the Tactile Internet ​
Author: Mubassir Serneabat Sudipto (Iowa State University), Shakil Ahmed (Grand Valley State University), Ashfaq Khokhar (Kansas State University)
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.NI
arXiv:2608.18572v2 Announce Type: replace-cross Abstract: Tactile Internet services couple cyber events directly to physical actuation, so security decisions must improve risk discrimination without perturbing the control path. This paper presents VQC-ZTI, a split-plane Variational Quantum Classifie...
244. FlavourBench: Executable Culinary Reward Maps for Language Model Evaluation and Post-Training ​
Author: Josef Chen (Independent Researcher), Erim Hayretci (Imperial College London)
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG, cs.SE
arXiv:2608.20574v2 Announce Type: replace-cross Abstract: Open-ended language-model evaluation often substitutes another model or a small preference panel for a missing answer key. We introduce FlavourBench, which instead compiles dense answer maps from a versioned culinary environment. Each task as...
245. Complexity Induction: Compositional Generalization via Structured Training Distortion ​
Author: Aleksandr V. Abramov
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.21464v2 Announce Type: replace-cross Abstract: We demonstrate that structured distortion of training data - which we term complexity induction - can induce compositional generalization in a standard CNN classifier without architectural modification. Using synthetic images of colored geome...
246. Keeping the Index Open: The Recommendation-Side Cost of Shared Search and Recommendation ​
Author: Theodore Rogers, Joe Standerfer, Dmitrii Timoshenko, Haoxue Li, Zuhaib Akhtar, Soyoung Yang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.24079v2 Announce Type: replace-cross Abstract: A shared search-and-recommendation index must score new items from features alone because search has no exploration slot. In a public log covering both surfaces over one catalog, $38.6%$ of held-out query-search impressions show an item neve...
247. Account Consistency from Gameplay Traces: Same-Player Verification in Counter-Strike 2 ​
Author: Xuchen Zhang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.24893v2 Announce Type: replace-cross Abstract: In competitive first-person shooter (FPS) games such as Counter-Strike 2 (CS2), account-integrity review often asks whether an account's recent behavior remains consistent with its historical operator. This consistency question arises in case...
248. Retrieved But Not Reliable: A Survey on Attacks, and Defenses in Retrieval-Augmented Generation ​
Author: Minh Tran, Cuong Dang, Tuc Nguyen, Khanh-Tung Tran, Minh Huynh Nguyen, Trinh Chau, Kien Le, Do Xuan Long, Jiahao Zhang, Fali Wang, Hoang D. Nguyen, Thanh Le, Suhang Wang
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.LG
arXiv:2608.24977v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) enhances large language models by grounding outputs in external knowledge, improving factuality and reducing hallucinations. At the same time, the retrieval-augmented pipeline introduces new robustness and...
249. Unsupervised Post-Training of Foundation Models: A Survey ​
Author: Yijie Xu, Qianyi Cai, Huizai Yao, Yili Wang, Tianfu Wang, Cehao Yang, Xingbo Yao, Zhiyu Guo, Aiwei Liu, Xuming Hu, Weiyu Guo, Hui Xiong
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CV, cs.LG, cs.MM
arXiv:2608.24982v2 Announce Type: replace-cross Abstract: Foundation-model post-training usually relies on human labels, preference data, stronger teachers, or executable verifiers. We study Unsupervised Post-Training (UPT): update-bearing adaptation on unlabeled inputs whose learning signal is deri...
250. DataKernelBench: Can LLMs Optimize Database Queries on GPUs? ​
Author: Gokul Karthik Kumar, Yotam Perlitz, Corey Lammie, Andrea Giovannini, Katja Hose
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.DB, cs.LG, cs.PL
arXiv:2608.25061v2 Announce Type: replace-cross Abstract: GPUs increasingly accelerate database systems, but query-specific peak performance still often relies on hand-written kernels. Existing LLM kernel benchmarks focus on machine learning operators, leaving irregular, heterogeneous, data-movement...
251. Learning New Facts with QLoRA: An Acquisition-Retention Frontier ​
Author: Estelle Zheng, S'ebastien Warichet, Emmanuel Helbert, Christophe Cerisara
Published: 8/28/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.25677v2 Announce Type: replace-cross Abstract: Parameter-efficient fine-tuning is often assumed to preserve pretrained capabilities because it updates only a small number of parameters. We show that this assumption depends strongly on adapter capacity. We study factual acquisition in a co...
252. LM-X: Explainable Action Modeling with Progress, Event, and Uncertainty Prediction for Generalist Robot Manipulation ​
Author: Jin Lou, Zhiyuan Jing, Andong Chen, Xupeng Wang, Yuan Xu, Yuexuan Li, Xingdong Zhu, Zhijie Zhu, Yingwei Ji, Wenpeng Nie, Yufei Liu, Boyang Xing, Lei Jiang, Yan Cui, Ying Chu, Jingxuan Zhu, Jingyi Li, Liangliang Chen, Jinyan Liu, Zhiqi Song, Jidong Zhang, Hongming Li, Yuchen Zhu
Published: 8/28/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.25757v2 Announce Type: replace-cross Abstract: Generalist vision--language--action (VLA) policies learn long-horizon behavior mainly through short-horizon action prediction and reveal little beyond sampled commands. This creates two coupled bottlenecks: a single action target must implici...