arXiv cs.LG - 2026-08-19 ​
555 items collected.
1. Learning Discrete Riemannian Metrics for Physical Fields with Cochain-Frame Equivarianc ​
Author: Dongzhe Zheng, Christine Allen-Blanchette
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14556v1 Announce Type: new Abstract: Physical fields on meshes require a separation between topology and geometry: conservation laws are topological and should be exact, while geometry, material response, and anisotropic coupling must be learned from data. Existing neural surrogates often...
2. Forward Pass Domain Adaptation (Without Cross-Layer Backpropagation) ​
Author: Rivaan Patil, Simon Dennis, Hao Guo, Kevin Shabahang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14563v1 Announce Type: new Abstract: Forward-Pass-Only MLP training (FPO) adapts large language models without a backward pass through the model body, achieving 2.7--3.2x the throughput of standard fine-tuning at ~40% less peak training memory, while leaving off-domain benchmarks within s...
3. Coarse-to-Fine Multi-Resolution Diffusion Models for Trajectory Generation in Urban Systems ​
Author: Wen Ye, Muyan Weng, Chuizheng Meng, Hao Niu, Yizhou Zhang, Yan Liu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14570v1 Announce Type: new Abstract: Understanding human mobility is critical for a wide range of urban applications, including traffic management, epidemic control, and urban planning. However, due to privacy concerns, the availability of large-scale public trajectory data remains limite...
4. Geometry Is Not Robustness: A Trajectory-Level Study of PGD Evaluation ​
Author: Dhairysheel Durgule
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14594v1 Announce Type: new Abstract: Projected Gradient Descent (PGD) is widely used to evaluate adversarial robustness, typically via final adversarial accuracy, which does not capture model behaviour throughout the attack. Recent work proposes trajectory-level diagnostics, such as loss ...
5. DumpsterCluster: From Dumpster Diving to Serving LLaMA-70B on $60 GPUs ​
Author: Zeyu Cao, Xuan Guo, Cheng Zhang, Cheuk Hang Lau, Ilia Shumailov, Yiren Zhao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR
arXiv:2608.14614v1 Announce Type: new Abstract: As AI datacenters retire functional GPUs, vast quantities of still capable accelerators enter secondary markets. This paper investigates whether these retired GPUs can find a productive afterlife to form a DumpsterCluster that can serve modern LLM infe...
6. Calibrated Trust, Not Sharper Prediction: An Empirical Test of Uncertainty Fusion ​
Author: Surya Saka
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.14617v1 Announce Type: new Abstract: A recurring proposal in legal AI is to improve case-outcome prediction by fusing uncertainty tools (evidence graphs with belief propagation, sequential Bayesian odds updating, Dempster-Shafer combination, and conformal prediction) into one pipeline. We...
7. PIKFNO: An Interpretable Neural Operator Based on Physics Informed Kernel Function ​
Author: Yuan Guo, Hanshu Chen, Zhuojia Fu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14619v1 Announce Type: new Abstract: This work proposes a new interpretable neural operator framework, termed the Physics Informed Kernel Function Neural Operator (PIKFNO), which explicitly incorporates physics informed kernel functions derived from governing equations into the neural ope...
8. Explaining Reinforcement Learning Decisions in Self-adaptive Systems ​
Author: Jasmina Gajcin, Juan C. Rosero, Ivana Dusparic
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14620v1 Announce Type: new Abstract: Reinforcement Learning (RL) has been extensively used in autonomous and self-* systems, but RL policies, especially deep RL ones relying on neural networks, lack transparency and are difficult to understand. This can lead to diminished user trust, and ...
9. Metaplasticity as adaptive gradient preconditioning for incremental learning ​
Author: Isabelle Aguilar, Zayn Andre Zainal, Omid Kavehei
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14634v1 Announce Type: new Abstract: Biological intelligence naturally prevents catastrophic forgetting through Complementary Learning Systems (CLS) theory, a macroscopic consolidation process driven at the local level by synaptic metaplasticity: the continuous, history-dependent neuromod...
10. Fractional Optimizers Meet Fractal Activation Functions: An Empirical Study of Multi-Scale Optimization in Neural Network ​
Author: Sebastian Raubitzek, Georg Goldenits, Sebastian Schrittwieser, Philip K"onig, Kevin Mallinger
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14636v1 Announce Type: new Abstract: Fractional optimization methods and fractal activation functions are two independent directions for improving neural network training. Fractional optimizers extend first-order optimization through fractional derivatives and memory effects, whereas frac...
11. Early Cycle Charge Trajectory Generative Prediction and Full Life Cycle Health Management of Iron-Chromium Flow Batteries Based on FlowBD-E1 ​
Author: Suyang Zhuang, Zekun Jiang, Tianhang Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ME
arXiv:2608.14637v1 Announce Type: new Abstract: Long-duration stationary energy storage requires batteries whose degradation can be detected before substantial capacity loss has accumulated. Iron-chromium redox flow batteries are attractive for this role because they use abundant and low-cost active...
12. Randomly initialized autoencoders: fixed points and edge-of-chaos ​
Author: Leonid Berlyand, Roman Sarapin, Yitzchak Shmalo, Victor Slavin, Sasha Sodin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math.PR, math.ST, stat.TH
arXiv:2608.14638v1 Announce Type: new Abstract: In this paper we study autoencoders, a special class of deep neural nets (DNNs) whose performance can be characterized via their fixed points. This perspective naturally raises questions of existence, stability, and basins of attraction of these fixed ...
13. Valid Per-Field Selective Risk Control for Document Extraction: Three Failure Modes, a Validity Ladder, and When Conditioning Pays ​
Author: Bhaskar Gurram
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.14639v1 Announce Type: new Abstract: Per-field accept/review with selective risk at most alpha -- accept a field only if the error rate among accepted fields is controlled -- is the trust contract document-extraction systems need, and the natural procedure silently violates it on real doc...
14. BDIP-Net: Dual-Interaction Graph Learning for Property Prediction of Bilayer Materials ​
Author: An Vuong, Chen Zhao, Jin Hu, Shui-Qing Yu, Xintao Wu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.AI
arXiv:2608.14640v1 Announce Type: new Abstract: Stacked bilayer materials exhibit rich stacking-dependent properties driven by the interplay between strong intra-layer bonding and weak inter-layer van der Waals interactions. The computational discovery of such materials is challenging because accura...
15. Training and Evaluating Ethical Reinforcement Learning Agents on Per-Episode Distributions ​
Author: Prabhjyot Singh, Majid Ghasemi, Mark Crowley
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14642v1 Announce Type: new Abstract: Reinforcement Learning (RL) agents trained on a single reward signal exploit the gap between the designed reward and the intended behavior. This is particularly a problem when we are trying to imbue ethical behavior into RL agents. An agent can look et...
16. DUET: Dual-Teacher On-Policy Distillation via Same-Weight Disagreement for Prohibition Compliance ​
Author: Zihan Li, Feifei Li, Wenhui Que
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.14644v1 Announce Type: new Abstract: Real-world LLM deployments increasingly rely on runtime-injected prohibitions--enterprise policies, PII redlines, tool boundaries--that vary per request and per tenant. Conventional post-training is structurally ill-suited: SFT hides the violation sign...
17. Efficient Neural-Network-Based High-Resolution Radiative Transfer for CO___ Retrieval, and Application to Interferometric Sensing ​
Author: Jordan Lontsi Tedongmo (CB), Yann Ferrec (CB, IFUMI), Laurence Croiz'e (CB, IFUMI), Pablo Mus'e (CB, IFUMI), Gabriele Facciolo (CB), Andr'es Almansa (MAP5 - UMR 8145, IFUMI)
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph
arXiv:2608.14645v1 Announce Type: new Abstract: Studying climate change requires reducing uncertainties in CO2 and CH4 emission estimates to better distinguish anthropogenic from natural sources, which motivates spaceborne measurements with improved revisit frequency and spatial coverage. In this co...
18. iFuzz-Meta: An Interpretable Fuzzy Learning Framework Bridging Top-Down and Bottom-Up Knowledge Integration ​
Author: Xiaowei Jiang, Daniel Leong, Beining Cao, Nan Zhou, Yingtao Ren, Yu-Cheng Chang, Thomas Do, Chin-Teng Lin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC
arXiv:2608.14646v1 Announce Type: new Abstract: Interpretable representation learning remains a key challenge in modern neural computation, particularly when models are expected not only to perform but also to explain their reasoning. This paper introduces iFuzz-Meta, an interpretable fuzzy rule-bas...
19. SMOPD: Selective Token-Entropy Masking for Dirty-History Multi-Turn On-Policy Self-Distillation ​
Author: Chenyang Jiang, Changhan Huang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14647v1 Announce Type: new Abstract: Dirty-history rollouts make multi-turn on-policy self-distillation (OPSD) brittle: once a student emits an erroneous intermediate reply, later turns are conditioned on that reply, and uniform distillation can spend loss on tokens that carry little corr...
20. Discrete Diffusion Language Models Are Training-Free Multi-Label Classifiers ​
Author: Pawan Kumar
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14649v1 Announce Type: new Abstract: We present dLLM-SetScore, a training-free method that uses discrete masked-diffusion language models for multi-label text classification. For each candidate label, it asks a short yes/no question and compares the probabilities of the two answer tokens ...
21. Paired Exact-Reset Evaluation of a Prediction-Derived Medium-to-Full World-Model Cascade ​
Author: Malo de Pastor
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2608.14650v1 Announce Type: new Abstract: Existing adaptive-inference and world-action-model systems use cheap-stage outputs or predicted futures to allocate additional computation. We study a narrower question: under paired exact-reset physical outcomes, can a Medium-derived interface predict...
22. Pushing the Limits of High-Resolution Weather Forecasting through Data Scaling ​
Author: Yang Zhao, Peisong Niu, Tian Zhou, Ziqing Ma, Guanlong Ma, Rong Jin, Huiling Yuan, Liang Sun
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.14652v1 Announce Type: new Abstract: The development of 0.1$^{\circ}$ global weather forecasting models based on machine learning (ML) is constrained by the limited availability of high-resolution data, as decades of reanalysis are only available at 0.25$^{\circ}$ resolution. While existi...
23. Do Uncertainty Signals Help? A Systematic Study of Uncertainty-Aware Decoding with Rollback Mechanisms ​
Author: Xianzong Wu, Xiaohong Li, Yuejun Guo, Xinyang Liu, Tianlin Li, Junjie Wang, Qiang Hu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14653v1 Announce Type: new Abstract: Prediction uncertainty is a widely adopted metric for quantifying model confidence, with downstream applications spanning model explanation, data selection, and prediction rollback. Despite its demonstrated utility, the potential of uncertainty quantif...
24. FedImp: Enhancing Federated Learning Convergence with Impurity-Based Weighting ​
Author: Hai Anh Tran, Cuong Ta, Truong X. Tran
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14654v1 Announce Type: new Abstract: Federated Learning (FL) is a collaborative paradigm that enables multiple devices to train a global model while preserving local data privacy. A major challenge in FL is the non-Independent and Identically Distributed (non-IID) nature of data across de...
25. Diagnosing and Mitigating Perception-Decision Misalignment in Omni-LLMs via Modality Subspace Activation ​
Author: Hongbo Jiang, Jie Li, Yunhang Shen, Tianyu Xie, Pingyang Dai
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.14655v1 Announce Type: new Abstract: Omni-Large Language Models (Omni-LLMs) power complex multi-modal reasoning in applications like World Action Models and autonomous agents. However, their strong performance often masks a profound Perceptual-Decision Misalignment (PDM), where decisions ...
26. P2E-VQ: ECG-linked representation augmentation for PPG via discrete patch retrieval ​
Author: Zhongli Wu, Zhuangzhi Gao, He Zhao, Feixiang Zhou, Fu Wang, Jinru Ding, Yuankai Wang, Hongyi Qin, Gregory Y. H. Lip, Bil Kirmani, Yalin Zheng
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14656v1 Announce Type: new Abstract: Photoplethysmography (PPG) is widely used in consumer wearables because of its low cost and ease of acquisition. However, unlike electrocardiography (ECG), PPG measures peripheral pulse dynamics rather than cardiac electrical activity, limiting its abi...
27. LUNG-KGMM: Knowledge-Guided Multimodal Learning for Lung Cancer Incidence Prediction ​
Author: Chunlei Yang, Shuyan Li, Zhong Cao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.14657v1 Announce Type: new Abstract: Early identification of lung cancer risk is critical for timely intervention, yet existing prediction models are limited by their reliance on single data modalities and their inability to leverage structured clinical knowledge. We propose LUNG-KGMM, a ...
28. pico-type: A 1.5M-Parameter Byte-Level Multi-Head Content Classifier ​
Author: Gautam Kishore
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.IR
arXiv:2608.14658v1 Announce Type: new Abstract: We introduce pico-type, a byte-level multi-head content classifier with approximately 1.5 million parameters that simultaneously predicts seven content properties from raw UTF-8 bytes in a single forward pass. Operating directly at the byte level -- no...
29. Ring-based Spatial Transformer: Learning Non-linear Spatial Interactions between Building Distribution and Pedestrian Flow ​
Author: Shun Nakayama, Takahiro Kanamori, Wanglin Yan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY
arXiv:2608.14660v1 Announce Type: new Abstract: This study proposes a ring-based SpatialTransformer to learn how building uses at different distances from a railway station interact to generate pedestrian flow. Concentric ring buffers at 100-meter intervals up to 800 meters were defined around 100 r...
30. An automatic-differentiation framework for time-lapse electrical resistivity tomography inversion of hydrologic dynamics ​
Author: Pu Yang, Zhengyang Fang, Yuxin Liu, Xuan Su, Deshan Feng, Hang Chen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14661v1 Announce Type: new Abstract: Time-lapse electrical resistivity tomography (TL-ERT) provides spatially distributed information on subsurface hydrologic changes. However, inversion of long monitoring sequences is computationally demanding. Modifying the data misfit, regularization, ...
31. In-Context Learning to Assess Built Environment Impacts on Perceived Neighborhood Walkability Among Mobility-impaired Older Adults ​
Author: Houhao Liang, Kresimir Friganovic, Joanne Kua, Noor Hafizah Ismail, Su Su, Bryan Yijia Tan, Navrag B. Singh, Panos Mavros
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.AP
arXiv:2608.14663v1 Announce Type: new Abstract: As global populations age, enhancing neighborhood walkability through inclusive urban design is important for mitigating built environment (BE) barriers that discourage physical activity and social participation among older adults. This study investiga...
32. Quantifying Depth Sufficiency in Residual Neural Networks: A First-Order Criterion ​
Author: Zeyu Liu, Jinhao Zhang, Yunquan Zhang, Guangming Tan, Xiang Gao, Fangming Liu, Daning Cheng
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14664v1 Announce Type: new Abstract: How can we determine whether a trained neural network is already deep enough? We study this under a fixed function-preserving residual-growth protocol specifying insertion locations, residual families, zero-output initializations, and zero-state first-...
33. When Does the Best Sampling Temperature Rise with the Budget? Sufficient Conditions for Pass@k ​
Author: Changsu Jeong (Independent Researcher)
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14665v1 Announce Type: new Abstract: The temperature that maximizes pass@$k$ is often low for a small sampling budget and higher for a large budget. This pattern has been reported from Codex through recent multi-sample inference studies. It is not an algebraic property of pass@$k$: as Slo...
34. ARGUS: Attention-Guided Transformers for Scalable Person Identification Using Wi-Fi Telemetry ​
Author: Nayan Sanjay Bhatia, Pranay Kocheta, Yuhan Li, Katia Obraczka
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.14670v1 Announce Type: new Abstract: Passive, device-free person identification offers an alternative to camera- and wearable-based biometrics, yet existing wireless approaches rely largely on gait or activity cues and are rarely evaluated at scale. In this paper, we present \emph{Argus},...
35. Take it Personally: The Limits of General SSL Representations for Real-Life PPG Emotion Detection ​
Author: Dominika Kunc, Przemys{\l}aw Kazienko, Stanis{\l}aw Saganowski
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14675v1 Announce Type: new Abstract: While Self-Supervised Learning (SSL) effectively extracts general representations from noisy, unconstrained physiological signals such as photoplethysmography (PPG), its suitability for highly subjective tasks remains unproven. In this work, we evaluat...
36. RouteTS: Frequency-Time Routing for Time Series Forecasting ​
Author: Gaofeng Lin, Lei Duan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.14682v1 Announce Type: new Abstract: Real-world time series inherently intertwine global periodic structures with localized non-stationary variations. Existing approaches process these heterogeneous dynamics within a single computational domain, incurring fundamental limitations: time-dom...
37. One Score, Two Decisions: Selective Prediction on the Rare-Disease Tail ​
Author: Zhaoyang Jiang, Zhizhong Fu, Yunsoo Kim, Zicheng Li, Xuanqi Peng, Fei Teng, Jiacong Mi, Honghan Wu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14683v1 Announce Type: new Abstract: Given a patient's clinical findings, a diagnostic system ranks possible diseases and must decide when to endorse its first prediction or defer it for review. This decision is usually made by thresholding the top score. Selective prediction over ranked ...
38. Mitigating Rubric Interference in LLM Judges via On-Policy Self-Distillation ​
Author: Dingyao Yu, Tong Zhang, Yutao Mou, Yunxiao Zhang, Wei Ye, Shikun Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14684v1 Announce Type: new Abstract: LLM judges increasingly evaluate responses against fine-grained rubric checklists. When a sample requires multiple rubrics, current methods typically assess each in a separate inference call. Evaluating all rubrics in a single pass is a natural alterna...
39. Rethinking Reverse KL as Adaptive Entropy Distillation ​
Author: Shizhen Li, Zhiyu Shen, Yuyin Lu, Yunhe Pang, Jielin Song, Yanghui Rao, Fu Lee Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.14685v1 Announce Type: new Abstract: Knowledge distillation (KD) is widely used to transfer the capabilities of large language models (LLMs) to smaller students, but existing objectives often struggle to balance faithful imitation and robust generation. In particular, existing methods mai...
40. A Reproducibility Study of Partial Residual Ablations in Pre-LN Transformers ​
Author: Pratikkumar Babariya
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14689v1 Announce Type: new Abstract: Residual connections are a fundamental component of transformer architectures, yet the roles of the attention and feed-forward residual pathways remain poorly understood when considered independently. This paper presents a reproducibility study of part...
41. The Quantum Shortcut: Complex Phase-State Dynamics Reduce the Optimization Steps of Sequence Models ​
Author: Ahmed Nebli, Hadi Saadatdoorabi, Christopher Keibel, Kevin Yam
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, quant-ph
arXiv:2608.14691v1 Announce Type: new Abstract: Sequence models are conventionally distinguished by their backbone, the mechanism that routes information across positions, such as attention or recurrence. This paper varies a choice that is prior to the backbone and shared by nearly all current model...
42. Tail-Aware Top-$k$ On-Policy Distillation ​
Author: Huipeng Huang, Hongxin Wei
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14728v1 Announce Type: new Abstract: On-policy distillation (OPD) has emerged as an effective paradigm for transferring knowledge between language models, where a student is trained to align its next-token distribution with the teacher's along its own trajectories. To provide dense superv...
43. A Novel Fourier Feature Network for Solving Partial Differential Equations ​
Author: Qihong Yang, Zhijie Su, Yangtao Deng, Qiaolin He
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14733v1 Announce Type: new Abstract: Building on the foundation of single-hidden-layer neural networks, Fourier Feature Networks (FENs) are proposed, which incorporate Fourier features using $\cos$, $\sin$, or a combination of both. Similar to Extreme Learning Machines (ELMs), FENs employ...
44. Unraveling the Size Determination Mechanism of Nanocrystal Synthesis via Interpretable Neural Networks ​
Author: Kai Gu, Haizheng Zhong
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.AI
arXiv:2608.14734v1 Announce Type: new Abstract: Deep learning models of nanocrystal synthesis enable the prediction of size and shape by encoding precursors and reaction conditions. However, their black-box nature hinders gaining deep insights into the underlying synthetic mechanisms. Here, we devel...
45. Generative Learning of Separatrices ​
Author: Ellis R. Crabtree, Dimitris G. Giovanis, Anastasia Georgiou, George Datseris, Ioannis G. Kevrekidis
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math.DS, stat.ML
arXiv:2608.14743v1 Announce Type: new Abstract: The identification and reconstruction of the boundaries separating basins of attraction in multistable, multidimensional dynamical systems presents a fundamental challenge in computational dynamics. These structures govern transition pathways and other...
46. Iterative Refinement Diffusion for Super-Resolved Data Assimilation of Multiscale Physical Systems ​
Author: Mrigank Dhingra, Ramchandran Muthukumar, Rebecca Willett, Omer San
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn
arXiv:2608.14744v1 Announce Type: new Abstract: Recovering high-resolution states from sparse, low-resolution observations is a central challenge in scientific machine learning and data assimilation. Classical data assimilation exploits temporal information through forecast-analysis cycles, but ofte...
47. WANDR: A Benchmark for Wide and Deep Research ​
Author: Vitaliy Polshkov, Marcin Pitera, Jeremy Yang, Kirill Priemko, Maksim Gaiduk, Aleksandr Nikolenko, Denis Bykov, Clare Southern, Denis Yarats, Jerry Ma
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14747v1 Announce Type: new Abstract: WANDR (Wide ANd Deep Research) is a benchmark of 500 realistic, challenging data-collection tasks for research agents. Each task requires a system to discover a large set of entities that satisfy specified criteria (breadth), investigate each entity th...
48. Real-Time State-of-Health Estimation and Online Degradation Prognosis from Partial Battery Discharge Using Physics-Informed Neural Networks ​
Author: Bego~na Ispizua, Serio Gil-L'opez, Leire Arrizabalaga, Ibai La~na
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14764v1 Announce Type: new Abstract: With the increasing integration of renewable energy sources, energy storage systems have become essential, making the accurate estimation of their State of Health (SOH) and degradation behavior critical. In this work, we propose a physics-informed deep...
49. ER-KANs: Efficient and Robust Kolmogorov-Arnold Networks for Data-Scarce Scientific Machine Learning ​
Author: Harshil Lodhiya
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14773v1 Announce Type: new Abstract: The efficient-KAN literature---covering Chebyshev, wavelet, and radial-basis-function variants of the original Kolmogorov-Arnold Network---has been benchmarked almost entirely on clean data. We show that this choice conceals a large capability differen...
50. p-Spin Glass Network Efficient Single-Batch Continual Learning ​
Author: Vladimer Khasia
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14774v1 Announce Type: new Abstract: Modern sequence models heavily rely on massive memory footprints and large-batch stochastic optimization, barriers that restrict sample efficiency and continual learning. We introduce the $p$-Spin Glass Network, a novel architecture that overcomes thes...
51. NRCD: An Open Database of Collegiate Running with Unified Performance Standardization ​
Author: Jonathan A. Karr Jr., Ryan M. Fryer, Ben Darden, Nicholas Pell, Kayla Ambrose, Evan Hall, Ramzi K. Bualuan, Nitesh V. Chawla
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CY, cs.IR
arXiv:2608.14776v1 Announce Type: new Abstract: Collegiate running in the United States generates thousands of race results annually in cross country and track and field, yet no large-scale dataset has been publicly available for research. Existing websites such as Athletic.net, MileSplit, and TFRRS...
52. Is Grokking a Loss of Normal Hyperbolicity of the Interpolation Manifold? ​
Author: Suvinava Basak
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14803v1 Announce Type: new Abstract: A recent line of work recasts the post-memorization phase of grokking as constrained optimization: once a network interpolates the training set, weight decay drives a slow drift along the zero-loss manifold toward lower norm. In the language of dynamic...
53. Disentangling Homophily and Rarity: Explaining Failure in Graph Neural Networks ​
Author: Preben M. Ness, Fariz Ikhwantri, Dusica Marijan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14823v1 Announce Type: new Abstract: Are heterophilic nodes in a graph harder to classify because they are heterophilic or because they are rare? Some existing work frames classification of such nodes as a subgroup generalisation problem, where a model performs well on the majority group ...
54. M-LINKX: Multiview Graph Learning for Brain Cognitive Disease Detection ​
Author: An Phan, Yufei Jin, Xingquan Zhu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14847v1 Announce Type: new Abstract: Electroencephalogram (EEG) is a non-invasive and relatively low-cost procedure that measures brain electricity for the detection of cognitive diseases. EEG-based classification of dementia-related conditions, including Alzheimer's disease (AD), mild co...
55. Can Neural Networks Learn by Experimenting on Themselves? Self-Interventional Learning from Functional Consequences to Predictive Self-Knowledge ​
Author: Micha{\l} Tomaszewski
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14894v1 Announce Type: new Abstract: Machine-learning systems usually model external data, while their internal functional organization is analyzed by external observers. This work introduces Self-Interventional Learning (SIL), in which a neural system perturbs its own functional structur...
56. Degeneracy Counting Quantum Algorithm using Decoherence ​
Author: Malay Marut Das, Mark A. Novotny, Yaroslav Koshka
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, quant-ph
arXiv:2608.14941v1 Announce Type: new Abstract: Counting the global optima of a classical optimization problem is a #P-hard task. We develop the canonical thermal pure quantum (CTPQ) state-based degeneracy counting (CTPQsd#) algorithm that determines the number of global optima of a classical optimi...
57. PathFinder: Joint Decompositions of Linked Multimodal Datasets ​
Author: Ying-Qiu Zheng, Alex Fung, Stephen M Smith, Rogier B Mars, Saad Jbabdi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, eess.IV, q-bio.QM, stat.ML
arXiv:2608.14951v1 Announce Type: new Abstract: Low-rank matrix decompositions can uncover patterns and structure in data and have a number of different applications across many disciplines. Extensions to "joint" low-rank decompositions have been proposed to link datasets from different modalities. ...
58. LLM-based Framework for Generating and Verifying Parallel DEVS Statecharts ​
Author: Vamsi Krishna Vasa, Hessam S. Sarjoughian, Edward J. Yellig
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.LO
arXiv:2608.14956v1 Announce Type: new Abstract: The development of models demands sound modeling and simulation knowledge as well as domain knowledge. Every model should accurately represent a system's dynamics and be verifiable. Toward this objective, this research introduces an agentic PDEVS-LLM f...
59. Command-Space Counterfactual Explanations for Pareto-Conditioned Reinforcement Learning ​
Author: Joanikij Chulev, Hendrik Baier
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC
arXiv:2608.14963v1 Announce Type: new Abstract: Pareto Conditioned Networks learn multiple multi-objective reinforcement learning behaviours by conditioning a single policy on a desired return command. However, the local mapping from command and state to action remains opaque. We propose command-spa...
60. A Physiology-Informed Digital Twin Framework for Simulating Liver Health Progression ​
Author: Sumaiya Afroz Mila, Sandip Ray
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.14969v1 Announce Type: new Abstract: We present a physiology-informed digital twin of the human liver designed for longitudinal simulation of liver function and early-stage disease progression. The model, referred to as HEPATWIN, integrates key hepatic processes, including carbohydrate, l...
61. Do Geometry-Aware Positional Encodings Help Transformers in Spatial Imperfect-Information Games? ​
Author: Wenji Fu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.14982v1 Announce Type: new Abstract: Transformers applied to spatial imperfect-information games must represent map geometry while tracking hidden entities through time. We ask whether geometry-aware positional encodings improve these capabilities, without claiming a new positional encodi...
62. Lipschitz Bandits with Arbitrary Feedback Delays ​
Author: Yuhao Liu, Yu Chen, Longbo Huang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15036v1 Announce Type: new Abstract: The Lipschitz bandit problem extends the traditional multi-armed bandit framework to continuous action spaces by assuming that the reward functions satisfy a Lipschitz condition. This work investigates Lipschitz bandits under arbitrary feedback delays,...
63. Certifying Compressed Language Models: An Audit and a Statistical Toolkit ​
Author: Amogh Singh
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15046v1 Announce Type: new Abstract: A fraction of a point of benchmark accuracy is the usual evidence that a compressed model is equivalent to its original. That quantity is least informative when two models are most alike: a net delta is what survives cancellation between opposing per-i...
64. Online Convex Optimization with Dueling Feedback ​
Author: Yiyang Lu, Hareshkumar Jadav, Mohammad Pedramfar, Ranveer Singh, Vaneet Aggarwal
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15050v1 Announce Type: new Abstract: We study online convex optimization with dueling (pairwise comparison) feedback, where the learner observes only a binary preference between two queried points. While dueling feedback is well understood in discrete or stochastic settings, the adversari...
65. A Unified Mamba--MoE Surrogate for Closed-Loop Simulation and Measurement-Window Forecasting of Inverter Transients ​
Author: Haoguang Wang, Huy Hoang Le, Akhila Kandivalasa, Christian Moya, Marcos Netto, Guang Lin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2608.15051v1 Announce Type: new Abstract: This paper proposes a Mamba surrogate model with mixture-of-experts (MoE) routing to represent the transient dynamics of inverter-based resources. A Mamba surrogate model is a predictive machine learning model built on the Mamba architecture. MoE routi...
66. GATTA: Graph Active Learning with Test-Time Augmentation ​
Author: Zsombor B'anfi, Andr'as G'ezsi, Andr'as Formanek
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15084v1 Announce Type: new Abstract: Test-time augmentation (TTA) has proven effective for improving model robustness and uncertainty estimation in computer vision, yet its application to graph-structured data remains largely unexplored. We introduce GATTA (Graph Active Learning with Test...
67. EMASAM: a Computationally Efficient Sharpness-Aware Minimization via EMA-Guided Perturbations ​
Author: Tanapat Ratchatorn, Masayuki Tanaka
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.15105v1 Announce Type: new Abstract: Recent progress in optimization research has highlighted the sharpness of the loss landscape as a key factor in narrowing the generalization gap. Motivated by this insight, Sharpness-Aware Minimization (SAM) was proposed as a training strategy that enh...
68. Global Federated Learning Strategies for Building Efficient Personalized Models ​
Author: Seongyoon Kim
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15107v1 Announce Type: new Abstract: Federated learning (FL) is a practical framework that can train models on distributed user data while guaranteeing data privacy; however, due to heterogeneity in which each user has a different data distribution, problems frequently arise where both gl...
69. Probability-Preserving Transformer for the Time-Dependent Schr\"odinger Equation ​
Author: Mushtaq Ali, Muzamil Tariq, Niaz Ali Khan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, quant-ph
arXiv:2608.15112v1 Announce Type: new Abstract: Solving the time-dependent Schr"odinger equation (TDSE) via traditional numerical methods is computationally intensive. Transformer models offer a compelling alternative, but standard implementations rely on soft constraints that cannot rigorously gua...
70. Decision-Driven Regularization: A Blended Model for Learning and Optimization ​
Author: Gar Goei Loke, Qinshen Tang, Yangge Xiao, Xun Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2608.15124v1 Announce Type: new Abstract: In contextual optimization, the decision-maker seeks optimal decisions to minimize a cost function, that varies based on observed features. This context is common in many business applications ranging from on-demand delivery and retail operations to po...
71. PureTD: Reinforcement Learning for Backgammon Money Games with No Evaluation-time Search ​
Author: Alexander L. Strehl
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15146v1 Announce Type: new Abstract: We revisit Tesauro's TD-Gammon for backgammon money games in the setting of no evaluation-time search. Both checker play and cube action (use of the doubling cube) are learned from scratch via self-play reinforcement learning (RL), with minimal hand-co...
72. FinFraudBench: A Heterogeneous Graph Benchmark for Financial Fraud Detection ​
Author: Yixuan Chen, Hongyu Zhan, Jie Sheng, Weiyu Han, Shuai Chen, Tianyi Zhang, Xiao Tan, Jun Xia
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15177v1 Announce Type: new Abstract: The increasing complexity of digital financial systems has reshaped financial fraud detection from isolated transaction classification into relational risk reasoning over interconnected financial entities. This shift has motivated graph-based fraud det...
73. MiNO: Cotangent-bundle propagator learning for PDEs ​
Author: Gnankan Landry Regis N'guessan, Bum Jun Kim
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.NA, math.NA
arXiv:2608.15187v1 Announce Type: new Abstract: Scientific machine learning for partial differential equations commonly targets solution fields, as in physics-informed neural networks, or solution maps, as in neural operators. We study a third target: the propagator itself, a phase and amplitude in ...
74. Structuring Semantic Embeddings for Principle Evaluation: A Prototype-Guided Contrastive Learning Approach ​
Author: Che Shen, Junwei Su, Lingpeng Kong, Chuan Wu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15224v1 Announce Type: new Abstract: Reliable post-hoc evaluation asks whether already generated text satisfies a target criterion after generation. In this paper we study a focused frozen-embedding setting using principle-evaluation proxy tasks: toxicity detection, fine-grained emotion c...
75. Learning reshapes power-law anisotropy in internal representations ​
Author: Asahi Nakamuta, Jun-nosuke Teramae
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.15239v1 Announce Type: new Abstract: Power-law anisotropy in internal representations has been observed across a wide range of biological and artificial neural systems, from state-of-the-art language models to the mouse cerebral cortex. This anisotropy is a key geometric property of high-...
76. Earth Observation Foundation Models for Terrestrial Ecohydrology: From Representation Learning to Process Inference ​
Author: Yi Yu, Jian Peng, Yucheng Lin, Trevor F. Keenan, Thomas F. A. Bishop
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, physics.bio-ph
arXiv:2608.15282v1 Announce Type: new Abstract: Earth observation foundation models (EOFMs) are emerging as reusable representation frameworks for data-driven retrieval, prediction and process modelling within ecohydrology, which integrate EO, meteorological forcing and process models to characteris...
77. No Task Fails Every Time: Why One-Shot Audits Are Structurally Blind to Agent Damage ​
Author: Shiven Khurdi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15286v1 Announce Type: new Abstract: We introduce AgentRelBench, an environment-agnostic reliability instrument that computes ground-truth, severity-priced damage from database state diffs across repeated runs, with no LLM in the measurement path, demonstrated on EnterpriseOps-Gym. Across...
78. MAPLE: MoE Adaptive Plug-and-play Layer-wise Expert allocation ​
Author: Lie Li, Wen Li, Junxiao Shen, Gusheng Hu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15299v1 Announce Type: new Abstract: Sparsely-activated Mixture-of-Experts (MoE) Transformers universally fix the same number of routed experts across all layers, a convention that ignores the well-documented heterogeneity in layer-wise redundancy. We demonstrate that this uniformity is s...
79. Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning ​
Author: Alexandre L. M. Levada
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, stat.ML
arXiv:2608.15313v1 Announce Type: new Abstract: In this paper, we propose SHOPCA (Shape Operator-based Principal Component Analysis), a novel method for unsupervised metric learning and dimensionality reduction that incorporates differential geometric information into the covariance structure of cla...
80. Spectral Rank Certification for Foundation Model Adapters ​
Author: Mohammed Ahnouch, Lotfi Elaachak
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ME
arXiv:2608.15351v1 Announce Type: new Abstract: Nominal LoRA rank is a design parameter; calibrated spectral evidence is a separate inferential quantity. This article develops a finite-sample framework for inferring effective rank structure in public foundation-model adapters. The theoretical core i...
81. SAPE: Sandwich Adapters for Parameter Efficiency in Large Language Model Fine-Tuning ​
Author: Mohammad Aref Jafari-Raddani, Morteza Mohajjel Kafshdooz
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15360v1 Announce Type: new Abstract: While Parameter-Efficient Fine-Tuning (PEFT) has substantially reduced the hardware cost of adapting Large Language Models (LLMs) by decreasing the number of trainable parameters, recent studies have sought to further improve PEFT through parameter sha...
82. Does 1/2-Tsallis-INF Also Work Well for Best-Arm Identification? ​
Author: Jingxin Zhan, Yuze Han, Zhihua Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15365v1 Announce Type: new Abstract: Regret minimization (RM) and best-arm identification (BAI) are two fundamental objectives in multi-armed bandits. Among regret-minimizing algorithms, $1/2$-Tsallis-INF is a canonical best-of-both-worlds FTRL algorithm: it achieves logarithmic pseudo-re...
83. Beyond Field Accuracy: Two-Axis Diagnosis of Inverse-PINN Parameter Error ​
Author: Yifan Zhang, Qian Tao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.15373v1 Announce Type: new Abstract: Inverse physics-informed neural networks (PINNs) can reconstruct a field accurately while returning an incorrect physical parameter. We introduce a two-axis post-training diagnosis that separates finite-sample resolution under a specified observation-a...
84. Every Expert Counts: ExactMoE for Memory-Efficient W4A16 Inference ​
Author: Amjad Saab
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15383v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) language models reduce arithmetic by activating only a small subset of experts per token, yet deployment still requires storing and moving the full expert bank. We present ExactMoE, an inference design that applies symme...
85. Look Before You Lift: Visual and Quantitative Diagnostics for Topological Deep Learning ​
Author: Mathilde Papillon, Guillermo Bern'ardez, 'Alvaro Ball'on Barreiro, Marco Montagna, R'emi Devaux, Antoine Jardin, Nina Miolane
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15388v1 Announce Type: new Abstract: Topological deep learning (TDL) methods rely on lifting raw data into higher-order discrete domains such as simplicial complexes, cell complexes, and hypergraphs. In practice, this lifting step is often treated as a black box: practitioners select a li...
86. Towards a theory of inference-time alignment with unknown rewards ​
Author: Steve Hanneke, Hongao Wang, Mingyue Xu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15402v1 Announce Type: new Abstract: Generative model alignment has received broad interest, and significant progress has been made in supervised fine-tuning and inference-time computation. Yet, alignment has remained poorly understood from a statistical learning perspective. We formulate...
87. FAST-DeepONet: Factor-Augmented Branch Representations for High-Dimensional PDE Inputs in the Small-Sample Regime ​
Author: Jiyong Kwon, Bongseok Kim, Guang Lin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15408v1 Announce Type: new Abstract: Deep operator networks can become statistically unstable when partial differential equation inputs are observed at thousands of strongly correlated sensors but only a small number of operator samples is available. We introduce FAST-DeepONet, a branch r...
88. Invariant Pretraining for Robust Code Representations ​
Author: Yifeng He, Yundi Xu, Christopher Castro Gaw Gonzalo, Zili Wang, Hao Chen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE
arXiv:2608.15412v1 Announce Type: new Abstract: Encoder-based code representation models remain widely deployed for discriminative tasks such as clone detection and code classification, where their small size and low inference cost are decisive. Their robustness, however, is fragile: under invariant...
89. SAGA: Structure-Attended Generative Action Embedding Model that encodes Multi-Surface User Action Sequences ​
Author: Tsz Fung Pang, Po Jen Chen, Nimish Ronghe, Farhad Farahani, Bo Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2608.15429v1 Announce Type: new Abstract: Prior embedding models for sequential recommendation typically operate within a homogeneous action space, limiting their ability to capture cross-surface behavioral signals spanning distinct behavioral domains. We present SAGA, a generative action embe...
90. Detecting Money Laundering in Rwandan Mobile Money: A Machine Learning Framework ​
Author: Emmanuel Nahimana, Ya'e Ulrich Gaba
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, q-fin.RM
arXiv:2608.15447v1 Announce Type: new Abstract: Mobile money has widened financial access across Sub-Saharan Africa and enlarged the surface for money-laundering and terrorism-financing (ML/TF) activity in ecosystems dominated by high-volume, low-value transactions. Rwanda is a case in point: severa...
91. Not All Attention Is Equal: A Quantitative Survey of the EEI Trade-off ​
Author: Aditya Singh
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15459v1 Announce Type: new Abstract: Attention mechanisms have driven machine learning for a decade, from neural machine translation to language models that do general-purpose reasoning. This survey covers four connected threads: their formulation for sequence-to-sequence tasks, adaptatio...
92. High-Dimensional Nonparametric Change-Point Detection via Low-Rank Degree-Three Density Projection ​
Author: Guoqing Zhang, Zhaixin Chen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15466v1 Announce Type: new Abstract: Distributional changes can be invisible to means and covariances yet appear in skewness, asymmetric interactions, or other third-order structure. We develop a nonparametric change-point method that retains every degree-at-most-three coefficient of a de...
93. Population Structure Analysis of an Inbred Population using Quantitative Shape Phenotyping from Stereo Retinal Photographs ​
Author: Li Tang, Michael D Abramoff
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, q-bio.QM
arXiv:2608.15471v1 Announce Type: new Abstract: The population structure of an inbred population of 781 people on Norfolk Island in the Pacific, 318 of which are descendants of the original Mutineers of the Bounty, is analyzed phenotypically using shape from stereo retinal fundus photographs. Three-...
94. Optimal Lower Bounds for Networked Information Aggregation ​
Author: Ambar Pal
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.15472v1 Announce Type: new Abstract: The problem of networked information aggregation, studied in Kearns et al. (2026), involves a group of learners situated on the vertices of a directed acyclic graph $G$, each learning a linear predictor $\widehat Y$ for a fixed random variable $Y$ give...
95. Measuring Structured Predictability in Neural Training Dynamics: A Cross-Regime Study ​
Author: Fanqi Wang, Weisheng Tang, Hairong Qi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15483v1 Announce Type: new Abstract: Modern deep networks are trained through long update trajectories, yet their temporal organization remains less systematically characterized than architectures, losses, or optimizers. We study short-horizon predictability as a measure of temporal redun...
96. QSMP: finding representative time series subsequences through Quick Shift+Matrix Profile ​
Author: Carlos H. Mendoza-Cardenas, Rogers F. Silva, Austin J. Brockmeier
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15492v1 Announce Type: new Abstract: Finding representative waveforms in long time series has scientific and practical value in many domains, as it enables summarization and visualization of large time series datasets, and downstream tasks like classification and forecasting. We present h...
97. PERO: Efficient Robust Post-Training Foundation Models for Encrypted Traffic Classification ​
Author: Wumei Du, Jiarong Wen, Kaiyu Zhang, Zi Yang, Yiqin Lv, Longfei Zhang, Dong Liang, Zheng Xie
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.15504v1 Announce Type: new Abstract: Encrypted traffic classification is vital for network security, yet real-world deployments are inherently sensitive to rare but high-loss errors such as misclassification of malicious traffic. The encrypted traffic foundation model, as a promising gene...
98. UniFed-VLM: Federated Instruction Tuning for Vision-Language Models with Multiple Heterogeneity ​
Author: Pengyu Wang, Baochen Xiong, Xiaoshan Yang, Yifan Xu, Zhang Qimeng, Haifeng Chen, Changsheng Xu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15516v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated strong performance in multimodal understanding and generation. However, fine-tuning of VLMs typically relies on centralized data, which raises privacy concerns in certain domains (e.g. healthcare). Federa...
99. Guaranteed Adaptive Modality Acquisition: When the Policy Chooses Its Own Calibration Group ​
Author: Melika Baghi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15520v1 Announce Type: new Abstract: A multimodal system may begin inference holding only some of its inputs and may acquire the rest at a cost. With adaptive acquisition, the policy determines which inputs are ultimately observed, so we state the guarantee conditional on that terminal in...
100. Spectral Saliency for Machine Unlearning ​
Author: Cedar Site Bai, Amber Yijia Zheng, Raymond A. Yeh, Brian Bullins
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15548v1 Announce Type: new Abstract: Machine unlearning (MU) aims to remove the influence of specific training data while preserving model utility. As the name suggests, MU can be viewed as the inverse of learning, using gradient-based updates to reduce the influence of a forget-set by co...
101. Amortised Post-Hoc Explanation with Exact Preservation for Dynamic Graph Anomaly Detectors ​
Author: Iyad Assaad Nekka, Hamida Seba, Walid Khaled Hidouci, Karima Amrouche
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15559v1 Announce Type: new Abstract: Anomaly detection in dynamic graphs underpins financial fraud analysis, intrusion detection, and platform integrity, where automated decisions require human-interpretable justifications. StrGNN, the strongest performer in recent benchmarks, produces no...
102. SchurQuant: Groupwise Discrete Optimization for Layer-Wise LLM Quantization ​
Author: Gunjun Lee, Sehwan Son, Younjoo Lee, Byungjun Kim, Jung Ho Ahn
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15567v1 Announce Type: new Abstract: Weight-only post-training quantization (PTQ) enables the deployment of large language models under tight memory budgets, but accuracy often collapses at 2-3 bits. Existing backpropagation-free PTQ optimizers have two limitations: group decisions ignore...
103. GraniKV: Asymmetric Granularity KV-Cache Paging for Multi-Agent Systems with Long Shared Prefix ​
Author: Jinhyun Jeon, Sungjoo Yoo
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15584v1 Announce Type: new Abstract: Production paged-serving engines apply uniform paging granularity to the KV cache, even though the two regions of a multi-agent workload have opposite storage requirements: a long shared prefix demands contiguity, while the per-request suffix demands f...
104. Quantum Models with Multi-Stage Training for Compositional Concept Generalization ​
Author: Mina Abbaszadeh, Matilda Karabina Moore, Mehrnoosh Sadrzadeh, Martha Lewis
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15601v1 Announce Type: new Abstract: Compositional Concept Generalization (CoCoGen), the ability to systematically recombine learned primitives in novel contexts, is a key challenge for multimodal learning. In this work, we provide a solution using a compositional model of meaning that se...
105. FluxBin: Flexible LUT-based Ultra-low-bit LLM Inference by Algorithm-Kernel Synergy ​
Author: Qingyao Yang, Runming Yang, He Xiao, Wendong Xu, Junyu Chen, Haobo Liu, Chenchen Ding, Ruihan Hu, Yik-Chung Wu, Ngai Wong
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15602v1 Announce Type: new Abstract: While binary quantization theoretically promises extreme compression and acceleration for Large Language Models (LLMs), existing research often overlooks the necessity of specialized hardware kernels, thus failing to unleash the full acceleration poten...
106. Benchmarking Quantum Machine Learning for Power-System Attack Detection: Evaluation Choices Decide the Outcome Before the Models Do ​
Author: Md Rezwanul Islam
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, stat.ML
arXiv:2608.15617v1 Announce Type: new Abstract: Machine-learning detectors for power-system cyberattacks are themselves attack surfaces, and quantum machine learning has been proposed for them. We benchmark fidelity-kernel SVMs and variational classifiers against six tuned classical models on public...
107. In Defense of OCTA: The Reconstruction-Utility Gap in OCT-to-OCTA Synthesis ​
Author: Michael Chertok, Alon Tiosano, Orly Gal-Or, Lior Kramarski, Einav Baharav Shlezinger, Irit Bahar, Lior Wolf
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15626v1 Announce Type: new Abstract: Optical coherence tomography angiography (OCTA) images retinal blood flow, giving capillary-perfusion and foveal-avascular-zone biomarkers that grade diabetic-retinopathy ischemia. Because OCTA hardware is less common than structural OCT, recent work s...
108. Sparse Prototype Code Underlies Classification and Prediction Across Modalities ​
Author: Yehonatan Avidan, Daniel D. Lee, Haim Sompolinsky
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, cs.AI, stat.ML
arXiv:2608.15632v1 Announce Type: new Abstract: Neural representations have become a central tool for studying the internal mechanisms of modern AI models, yet their complex high-dimensional structure makes them difficult to interpret. We show that classification tasks give rise to a universal repre...
109. Generalised Transportability via Causal Abstractions ​
Author: Yorgos Felekis, Paris Giampouras, Fabio Massimo Zennaro, Theodoros Damoulas
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15645v1 Announce Type: new Abstract: Transporting a causal conclusion from a source study population to a target one is a fundamental problem in causal inference. The theory of transportability provides a criterion for when this is possible: given experimental data from the source and obs...
110. Sequential Multimodal Evidence Optimization for Product Media Ranking in E-Commerce ​
Author: Prasenjit Dey, Frank McIntyre, Arnab Sinha
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15662v1 Announce Type: new Abstract: On modern e-commerce stores, customers consume ordered slates of heterogeneous product media, such as images, videos, and 3D renders, before making purchase decisions. Existing media-ranking systems often optimize myopic engagement proxies such as clic...
111. SubZero+: Efficient Zeroth-Order LLM Fine-Tuning via Large Learning Rates ​
Author: Ziming Yu, Shuyao Xiao, Xingyu Zhao, Sike Wang, Pan Zhou, Peiyu Zang, Xiangda Yan, Yongjie Yang, Jia Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15665v1 Announce Type: new Abstract: Zeroth-order (ZO) optimization enables backpropagation-free fine-tuning of large language models, but existing ZO methods suffer from high-variance gradient estimators, making convergence unstable and highly sensitive to learning rates. We propose SubZ...
112. Large Discovery Models: Empirically-grounded Model-Based Open-Ended Search ​
Author: Zhongwei Yu, Yan Song, Xue Yan, Anjie Liu, Xingyu Lu, Yihang Chen, Huichi Zhou, Siyuan Guo, Luoyang Sun, Sihan Chen, Xiangning Yu, Jun Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15669v1 Announce Type: new Abstract: Scientific discovery often involves optimising expensive-to-evaluate objectives over vast, structured, and open-ended hypothesis spaces, such as molecules, protein sequences, and computer programs. Generative models such as large language models (LLMs)...
113. PL-Guard: Probabilistic Logic Reasoning for LLM Guardrails ​
Author: Satchit Chatterji, Shihan Wang, Giovanni Sileno, Erman Acar
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15673v1 Announce Type: new Abstract: Large language model guardrails can be viewed as policy-consistency problems: a system must determine which policy-relevant facts hold in a prompt-response pair and what those facts imply under a given policy. Common approaches, including policy prompt...
114. Learning Auditable Classifier Models: Source-Disjoint Tree Ensembles ​
Author: Srikumar Krishnamoorthy
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15725v1 Announce Type: new Abstract: Predictive models in clinical and regulated settings must be accurate and fully auditable. Tree ensembles deliver strong accuracy on tabular data, but their sequential boosting couples structure discovery with coefficient estimation, making compact per...
115. FirstDiff: One-Step Diffusion-Based Anomaly Detection for Multivariate Time Series via Initial Noise Prediction ​
Author: Ali Boudaghi, Alireza Nemati, Hadi Zare
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.15727v1 Announce Type: new Abstract: Diffusion models have recently shown strong potential for multivariate time-series anomaly detection by learning the distribution of normal data through iterative denoising. Existing diffusion-based approaches, however, typically perform anomaly detect...
116. TinyCast: Probabilistic Zero-Shot Forecasting with Computed Periodicity ​
Author: Armin Steinhauser
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15767v1 Announce Type: new Abstract: We introduce TinyCast, an attention-free zero-shot forecaster that emits a predictive distribution from 146,505 parameters, on the premise that at this size the periodic structure of a context is worth computing rather than learning. A zero-parameter s...
117. Temporal Graph Prototype-conditioned Conformal Prediction for Fraud Detection ​
Author: Xudong Chen, Shengbo Gong, Lu Cheng, Wei Jin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15768v1 Announce Type: new Abstract: Conformal prediction (CP) provides distribution-free coverage guarantees and has emerged as a principled tool for uncertainty quantification. In edge-level fraud detection on temporal interaction graphs, where false positives and false negatives both c...
118. Learning Stock Trading Policies via Barycenter-Based Adversarial Inverse Reinforcement Learning ​
Author: Arishi Orra, Himanshu Choudhary, Manoj Thakur
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.15770v1 Announce Type: new Abstract: Designing effective trading strategies using reinforcement learning remains challenging due to delayed and noisy rewards, poor exploration, and the difficulty of enforcing explicit risk constraints. In this work, we propose BRaG, a barycenter-based adv...
119. Routing Divergence Is Not Evidence of Behavioral Influence in Same-Weight MoE Self-Distillation ​
Author: Cedric Caruzzo, Donggeun Yoo, Tae Soo Kim
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.15787v1 Announce Type: new Abstract: Two Mixture-of-Experts (MoE) forward passes can share every weight yet route the same token through different experts. This creates a possible blind spot in same-weight self-distillation, where a demonstration-conditioned teacher supervises a query-onl...
120. CrevasseSeg: A Label-Efficient UAV Crevasse Segmentation Framework ​
Author: Steven Wallace, William D. Harcourt, Richard Hann, Aiden Durrant, Somayajulu Sripada, Georgios Leontidis
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15790v2 Announce Type: new Abstract: Crevasse mapping from uncrewed aerial vehicle (UAV) imagery matters for glaciological research and for field safety in glaciated terrain. Yet, pixel-level annotation of glacier surfaces is costly and requires domain experts. We introduce CrevasseSeg, a...
121. Cross-Entropy Risk Estimation for Language Models: Inconsistency Must Be Dense, and the Holdout Method Is No Exception ​
Author: Hanti Lin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.15798v1 Announce Type: new Abstract: Language models are compared by their held-out per-token cross-entropy risk---the quantity scaling laws are fitted to. We show that it cannot be consistently estimated. Consistency, or convergence to the estimand, is defined relative to a \emph{possibl...
122. A Pre-Specified Construction-Confirmation Test of Operation-Level Causal Transfer Across Finite Isomorphic Symbolic Domains ​
Author: Xinyi Shan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15809v1 Announce Type: new Abstract: Behavioral accuracy, linear decodability, and successful activation interventions do not by themselves show that a model carries an operation-level structure from one symbolic domain to another. We ask a narrower question in finite isomorphic state spa...
123. KOALA: Koopman Operator Learning for WiFi-Based Anticipatory Hum ​
Author: Quang-Anh N. D., Duc Pham Minh, Thao Phuong Pham, Minh Anh Nguyen, Huan X. Nguyen, Tuan Dang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15815v1 Announce Type: new Abstract: WiFi Channel State Information (CSI) has emerged as a privacy-preserving alternative to cameras for human pose estimation. However, existing approaches treat pose inference as an instantaneous regression problem and do not model temporal dynamics, maki...
124. Second-Moment Memory in Coordinatewise Adam ​
Author: Jeonseong Kim
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15824v1 Announce Type: new Abstract: Adam retains a moving average of past squared gradients in its denominator, but the optimization cost of this memory is not well understood. We show that second-moment memory can itself suppress progress toward the optimum even under finite-variance st...
125. Self-Supervised Auxiliary Task Discovery for Stable Reinforcement Learning in Stock Trading ​
Author: Arishi Orra, Himanshu Choudhary, Manoj Thakur
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP, stat.ML
arXiv:2608.15841v1 Announce Type: new Abstract: Reinforcement learning has gained increasing attention as a data-driven approach for stock trading. However, learning a policy that is both profitable and stable remains challenging due to non-stationary market behaviour and noisy reward signals. Auxil...
126. Geometry of Forgetting: Representation Flux in Continual Learning ​
Author: Maksim A. Kazanskii
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.15854v1 Announce Type: new Abstract: Catastrophic forgetting remains a fundamental obstacle to continual learning, where neural networks lose previously acquired knowledge while learning new tasks. Existing methods primarily mitigate forgetting through parameter regularization or experien...
127. TransfHAR: Self-Supervised Wrist Representations for On-Demand Activity Recognition ​
Author: Aidan Bradshaw, Riku Arakawa, Xin Liu, Karan Ahuja
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15861v1 Announce Type: new Abstract: Fine-grained wrist activity recognition can support applications such as procedural step guidance and context-aware assistance, yet acquiring labeled data for every new task, user, and activity granularity remains a bottleneck. We present TransfHAR, a ...
128. Feasible and Novel Synthetic Population Generation with Tabular and Sequential Travel Attributes ​
Author: Farbod Abbasi, Zachary Patterson, Bilal Farooq
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15867v1 Announce Type: new Abstract: Synthetic populations are critical inputs for activity-based travel demand models, yet generating realistic populations from limited survey data remains challenging. Small samples miss valid attribute combinations, known as sampling zeros, and generati...
129. Deploying Frontier Agentic Technology in MOOSEnger, a Multiphysics-Capable AI Assistant ​
Author: Zaid Abulawi, Mengnan Li, Guillaume Giudicelli, Yang Liu, Cody Permann
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2608.15881v1 Announce Type: new Abstract: The Multiphysics Object-Oriented Simulation Environment (MOOSE) is an open-source finite-element framework for building multiphysics simulation applications. Using a multiphysics environment effectively demands specialized expertise, creating a barrier...
130. Layers Matter: Why Continual Learning Regularization Should Be Layer-Adaptive ​
Author: Brian B. Moser, Ahmed Anwar, Tobias Christian Nauen, Shishir Muralidhara, Federico Raue, Ren'e Schuster, Stanislav Frolov, Andreas Dengel
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15901v1 Announce Type: new Abstract: Continual learning regularizers like EWC fight forgetting by penalizing changes from previous-task parameters with per-parameter importance, typically diagonal Fisher values. Per-parameter looks more flexible than per-layer, but each layer's diagonal F...
131. Information Geometry of Message Passing ​
Author: Mykola Lukashchuk, Kyrylo Yemets, Alex Ledbetter, .{I}smail \c{S}en"oz
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15922v1 Announce Type: new Abstract: We show that the natural-gradient stationary condition of variational inference has an edge-local form on a Forney-style factor graph. We start from the Bethe free energy and constrain a selected edge marginal to an exponential family. At a stationary ...
132. ReliaGate: Reliability Routing for Low-Stakes Wearable Stress Prediction ​
Author: Jaden Moon, Yu Wu, Arvind Pillai, Andrew Campbell
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.HC
arXiv:2608.15951v1 Announce Type: new Abstract: We study when a wearable stress system should surface a prediction rather than change it. In low-stakes reflection and summary settings, aggregate accuracy is insufficient because withholding can reduce error while leaving some people with little or no...
133. Beat the Counter First: A Baseline for Temporal-Graph Anomaly Detectors ​
Author: Omair Shafi Ahmed, Zohair Shafi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15965v1 Announce Type: new Abstract: Progress in streaming, edge-level graph anomaly detection (GAD) has been marked by increasingly elaborate architectures, from count-min-sketch chi square tests to memory-augmented attention networks. Yet the empirical gains attributable to this added c...
134. A Banach-Space Theory of Markovian Halpern Iteration for Non-Expansive Maps ​
Author: Ege C. Kaya, Arda Fazla, M. Berk Sahin, Abolfazl Hashemi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2608.15966v1 Announce Type: new Abstract: We study stochastic approximation of fixed points of a non-expansive operator when the oracle samples originate from a continuing Markovian trajectory. A direct block-minibatch implementation of Halpern iteration attains an expected last-iterate residu...
135. The Limits of Binding in Dual Encoders ​
Author: Kin Ian Lo
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CV
arXiv:2608.15971v1 Announce Type: new Abstract: Dual-encoder models such as CLIP score an image-caption pair by a single inner product of two independently computed unit vectors, and fail at binding, often scoring near chance when asked to distinguish "a red car and a blue dog" from "a blue car and ...
136. Fiber Fingerprints of Hidden Learning-State Dynamics ​
Author: Qinyou Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15976v1 Announce Type: new Abstract: A learning system can occupy execution states that are indistinguishable under every declared present-behavior readout yet respond differently to future training. We formalize this through fiber fingerprints: controlled future-learning response laws re...
137. Operator-Theoretic Generalization Bounds for Multitask Deep Learning ​
Author: Mahdi Mohammadigohari, Thomas Borsani, Giuseppe Di Fatta
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15982v1 Announce Type: new Abstract: We develop operator-theoretic generalization bounds for deep multi-output function classes by representing network layers as Koopman composition operators on vector-valued reproducing kernel Hilbert spaces. In vector-valued Sobolev RKHSs, we derive Rad...
138. Toward Optimal Second-Order Path-Length Guarantee for Adversarial Multi-Armed Bandits ​
Author: Mengxiao Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.15996v1 Announce Type: new Abstract: We study second-order path-length regret in adversarial $K$-armed bandits against oblivious loss sequences. Bubeck et al. [2019] designed an algorithm that achieves $\widetilde{\mathcal{O}}(K+\sqrt{KQ_{\infty,1}})$ regret, where $Q_{\infty,1}$ is the f...
139. Retrieval-guided Twin Fusion with Similarity-aware Contrast for Molecule-Text Alignment ​
Author: Shunshun Gu, Shengqi Qiu, Hang Zhou, Xiao Luo
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16005v1 Announce Type: new Abstract: This paper studies the problem of molecule-text alignment, which aims to project molecules and their textual descriptions into a joint latent space for downstream tasks including molecule search and molecular property prediction. Previous approaches ty...
140. Breaking the Compression Barrier: Cross-Architecture Compression Boundary Learning via Reverse Regrowth ​
Author: Zhaocen Liu, Satvik Praveen, Yi Sheng
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.16010v1 Announce Type: new Abstract: Model compression is critical for deploying networks on resource-constrained edge devices. While pruning-based methods can significantly reduce model size, they often suffer from abrupt performance collapse beyond a sparsity thresh-old, making it diffi...
141. RagGAD: Rationale-Aware Conditional Gaussian Mixture Normalizing Flow for Unsupervised Graph Anomaly Detection ​
Author: Junxin Lu, Jing Zhao, Shiliang Sun
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16018v1 Announce Type: new Abstract: Graph anomaly detection aims to identify nodes that deviate from normal behavioral patterns within graphs. However, existing methods largely rely on the homophily assumption, which makes it difficult to distinguish spurious affinities and to capture th...
142. Group ICA 2.0: Closing the Gap Between Subjects and Group Latent Decomposition with Copula-Linked Group ICA (CoLiG-ICA) ​
Author: Oktay Agcaoglu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ME
arXiv:2608.16029v1 Announce Type: new Abstract: Group Independent Component Analysis (gICA) is widely used to decompose high-dimensional functional MRI data into interpretable brain networks. However, conventional gICA primarily identifies components shared across subjects. This group-level assumpti...
143. AdROD: HyperNetwork-based Adversarially Robust Object Detection for Autonomous Driving ​
Author: Yuting Wu, Dongfang Guo, Xiangzhong Luo, Qun Song, Rui Tan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.16031v1 Announce Type: new Abstract: Camera-based object detectors are vulnerable to physical adversarial attacks designed to suppress detections. While adversarial training and input purification offer some protection, they often overfit to specific attack distributions and fail on adapt...
144. NICE: Scale-Stable Perturbations for Graph Neural Network Explanations via Noise Corruption ​
Author: Ziluowen Luo, Jun Yin, Ruochen Liu, Ming Cheng, Shirui Pan, Chengqi Zhang, Senzhang Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16038v2 Announce Type: new Abstract: Post-hoc Graph Neural Network (GNN) explainers commonly follow a Perturb-Query paradigm, inferring the importance of graph elements based on queried predictions to perturbed inputs. However, such perturbations often introduce substantial distribution s...
145. OceanLight: Efficient Global Ocean Forecasting via Geometry-Adaptive Unstructured Mesh Representation ​
Author: Wei Wu, Xiang Wang, Hongze Leng, Qingye Min, Junxing Zhu, Junqiang Song
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16070v1 Announce Type: new Abstract: Reliable global ocean forecasting is critical for climate monitoring, marine navigation, and extreme event early warning. Physics-based ocean forecasting models impose prohibitive computational costs, while existing deep learning approaches predominant...
146. Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization ​
Author: Yixuan Wang, Yifei Chen, Haichao Zhang, Haozheng Luo, Xander Wu, Jie Ni, Yun Fu, Nuno Vasconcelos, Yijiang Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16072v1 Announce Type: new Abstract: Reinforcement learning (RL) with group-relative advantages has become the de facto standard for post-training language model reasoners. However, when optimizing multiple reward objectives, existing methods typically scalarize the reward vector with a f...
147. DeepOHeat-v2: Self-Improving Operator Learning for Fast and Trustworthy Thermal Optimization in 3D-IC Design ​
Author: Xinling Yu, Yixing Li, Ziyue Liu, Xin Ai, Zhiyu Zeng, Hai Li, Zheng Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, physics.data-an
arXiv:2608.16080v1 Announce Type: new Abstract: Thermal-aware optimization of multi-die 3D integrated circuits evaluates many designs, each a costly heat-equation solve. Operator-learning surrogates replace this solve with a fast forward pass, ideally trained from physics alone, without labeled data...
148. Towards Reasonable Molecular Structure Elucidation from Infrared Spectroscopy with Chemical Feedback ​
Author: Yusen Tan, Hongyu Zhan, Hai-tao Yu, Changxi Chi, Wenjie Du, Jun Xia
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16082v1 Announce Type: new Abstract: Infrared (IR) spectra provide characteristic signals of molecular structure, which are often interpreted by experts via functional-group identification or library matching, making the process time-consuming and ambiguous. Recent machine learning method...
149. Behaviour Is an Incomplete Measure of Reasoning Development: Cross-surface pre-arrival accessibility and the limits of developmental inference in a recurrent-depth reasoner ​
Author: Simon Lam-Muir
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16085v1 Announce Type: new Abstract: Capability development is routinely inferred from behavioural thresholds, from final checkpoints, or from what a decoder can read out of a hidden state. These quantities need not identify the same event. We study a 30M-parameter recurrent-depth relatio...
150. Unifying Graph Neural Networks Through a Common Layer Equation ​
Author: Sai Karthik Navuluru, Siddhartha Shankar Das, Bo Ni, Hongjie Chen, Yu Wang, Baris Coskunuzer, Nesreen K. Ahmed, Franck Dernoncourt, Mahantesh Halappanavar, Tyler Derr, Ryan A. Rossi, Lakshman Tamil
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16097v1 Announce Type: new Abstract: Graph neural networks are commonly described through family-specific equations whose notation obscures shared computations and structural differences. We introduce a common layer equation that represents covered architectures through seven components: ...
151. AsyTO: Asymmetric Temporal Operator for Parameter-Efficient Multivariate Time Series Forecasting ​
Author: Xiachong Lin, Du Yin, Hao Xue, Wen Hu, Imran Razzak, Arian Prabowo, Matthew Amos, Flora D. Salim
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16098v1 Announce Type: new Abstract: Multivariate time-series forecasting faces a structural dilemma: sharing one temporal predictor across variables is parameter-efficient but forces heterogeneous variables through an identical history-to-future map, whereas learning an independent predi...
152. RetroMPA: A Molecular Property-Aware Auxiliary Framework for Enhancing Retrosynthesis Prediction ​
Author: Mianzhi Liu, Fan Xiao, Zhiliang Yu, Huayang Huang, Yuke Li, Yi Yang, Wenbo Liu, Yu Wu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16111v1 Announce Type: new Abstract: Retrosynthesis is a cornerstone of drug discovery and organic synthesis. While data-driven deep learning models have shown remarkable progress, they autonomously learn reaction patterns from extensive datasets with limited integration of established ch...
153. Multi-Feature Riemannian Hypergraph for Online Test-Time Adaptation of Motor Imagery Brain-Computer Interface ​
Author: Siqi Li (Peking University, Chinese Institute for Brain Research, Beijing), Zhi Li (NeuCyber Neurotech), Tong Liu (NeuCyber Neurotech), Shuai Zhang (NeuCyber Neurotech), Yanfei Jia (Beijing Medical University), Zhiqiang Yi (Beijing Medical University), Jue Xie (NeuCyber Neurotech), Ni Ji (Chinese Academy of Medical Sciences & Peking Union Medical College, Chinese Institute for Brain Research, Beijing)
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, eess.SP, q-bio.NC
arXiv:2608.16134v1 Announce Type: new Abstract: In clinical motor imagery brain-computer interface (MI-BCI) decoding, cross-day transferability and online operation remain two critical challenges. Hypergraphs can improve transferability by capturing higher-order sample relationships, yet existing hy...
154. REFLEX: Reflexive Equilibrium Fixed-point Learning for Endogenous eXchanges ​
Author: Vignesh Nagarajan, Shriraghav Ashok
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.GT
arXiv:2608.16155v1 Announce Type: new Abstract: In over-the-counter corporate bond markets, dealers compete for client trades by quoting bid and ask prices. Tighter quotes attract more business, but also informed customers more likely to trade ahead of adverse price moves, leaving the dealer holding...
155. A Tree-Structured Approach for Phishing Template and Attacker Attribution Analysis ​
Author: Unai Agirre, Imanol Jerico, Felipe Casta~no, Andrea Venturi, Francesco Zola
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET
arXiv:2608.16158v1 Announce Type: new Abstract: Phishing remains a persistent and evolving cybersecurity threat, with attack volumes reaching record levels. This growth is driven by the industrialization of phishing through widely available phishing kits and reusable templates, which enable cybercri...
156. Demystifying Oversmoothing in Sheaf Neural Networks: An Index-Theoretic Criterion ​
Author: Junwen Dong, Yuhan Peng, Hao Li, Huitao Feng, Kelin Xia
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math.DG
arXiv:2608.16180v1 Announce Type: new Abstract: To combat oversmoothing in Graph Convolutional Networks, Sheaf Neural Networks (SNNs) were proposed as a generalization by equipping the graph with a sheaf structure and replacing the graph Laplacian with a sheaf Laplacian $\mathcal{L}$. Existing analy...
157. Understanding and Stabilizing Deep Q-Learning via Controlled Bootstrapping and Regulated Value Dynamics ​
Author: Bozhou Chen, Yongyi Wang, Hanyu Liu, Xionghui Yang, Wenxin Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16182v1 Announce Type: new Abstract: Deep Q-learning (DQL) has achieved remarkable empirical success in reinforcement learning, yet its training process remains notoriously unstable. Existing studies often attribute instability to isolated factors such as overestimation bias or representa...
158. Multi-Granularity Sentiment Integration for LLM-Based Multimodal Sentiment Analysis ​
Author: Shanshan Lin, Yuesheng Wu, Chao Chen, Yizhe Yang, Zhihao Chen, Zexian Yang, Xiangwen Liao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16201v1 Announce Type: new Abstract: Multimodal sentiment analysis (MSA) aims to predict sentiment polarity and intensity from heterogeneous inputs such as text, audio, and vision. While large language models (LLMs) offer strong semantic priors for MSA, effectively incorporating audio and...
159. Conditional Evaluation of Language Models with Cheap Auxiliary Signals ​
Author: Zhi Zhang, Lingfeng Lyu, Yue Kang, Doudou Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.16210v1 Announce Type: new Abstract: Aggregate accuracy hides where models succeed and fail. Estimating conditional performance profiles from gold labels alone is expensive, while cheap auxiliary signals such as LLM-judge scores, pairwise comparisons, confidence scores, and judge-disagree...
160. Quantifying the Gap Between Laboratory Battery Test Patterns and Field Duty Profiles ​
Author: Chunyang Zhao, Chresten Tr{\ae}holt
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16212v1 Announce Type: new Abstract: Laboratory battery tests provide the main empirical basis for battery performance and degradation studies, but their operating patterns do not directly represent field duty profiles. This paper quantifies the gap by comparing six accessible evidence so...
161. Beyond Peak Backlog: Conditional Energy and Temporal Geometry in Capacity-Constrained Delayed Bandit Optimization ​
Author: Anling Xiang, Yuwen Yang, Yang Shen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16216v1 Announce Type: new Abstract: What is the right delay complexity when a learner can track only $C$ pending feedback items and discarded feedback is permanently lost? Existing one-point bandit convex optimization guarantees in this model pay $\sqrt{T\sigma_{\max}}$, where $\sigma_{...
162. A Privacy Study of Sparse Collaborative Inference ​
Author: Maximilian Andreas Hoefler, Karsten Mueller, Wojciech Samek
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16236v1 Announce Type: new Abstract: Collaborative inference (CI) splits a model between an edge device and a server, whereby the client computes an intermediate activation, transmits it, and the server completes the computation. This raises two concerns, the communication cost of the tra...
163. Optimizing Multi-Market Participation of Battery and Electrolyser Systems Based on Field Performance ​
Author: Chunyang Zhao, Stoyan Trenchev, Shi You, Chresten Tr{\ae}holt
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16238v1 Announce Type: new Abstract: The increasing share of renewable energy in power systems creates a need for fast-response and flexible resources to maintain system stability. With the expansion of electricity markets and ancillary service products, opportunities arise to stack reven...
164. The Trade-off Between Covariate Dependence and Latent Structure in Representation Learning ​
Author: Ma{\l}gorzata {\L}az\k{e}cka, Ewa Szczurek
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16245v1 Announce Type: new Abstract: Disentangled representation learning seeks latent representations whose indicidual dimensions each align with a distinct covariate. Unsupervised approaches typically target latent dimension independence, yet this gives no guarantee that the resulting d...
165. SAUL: Sharpness-Aware Augmented-Lagrangian Unlearning ​
Author: Jaewan Choi, Junyoung Yang, Sangdon Park
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16249v1 Announce Type: new Abstract: Machine unlearning in Large Language Models (LLMs) faces a critical trade-off between erasing target knowledge and preserving general utility. We propose SAUL (Sharpness-Aware Augmented-Lagrangian Unlearning), which formulates unlearning as a constrain...
166. Efficient Coreset Selection via K-Nearest Neighbor Graphs ​
Author: Yingfan Liu, Leiyu Zhang, Jiadong Xie, Mingzhe Wang, Jeffrey Xu Yu, Jiangtao Cui
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16270v1 Announce Type: new Abstract: Coreset selection reduces the cost of model training by replacing a large training set with a small representative subset. Existing gradient-approximation coreset methods such as CRAIG and cluster-based variants can preserve model accuracy. Still, thei...
167. Foresight-England: Development of a National-Scale Generative AI Model of Electronic Health Records for Medical Event Prediction across the COVID-19 Pandemic ​
Author: Simon Ellershaw, Christopher Tomlinson, Zeljko Kraljevic, Spiros Denaxas, Harry Hemingway, Cathie Sudlow, Angela M. Wood, Anoop D. Shah, Richard Dobson
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16273v1 Announce Type: new Abstract: Foresight-England (Foresight-E) is the first national-scale generative foundation model of electronic health records (EHRs), developed as a research pilot strictly for COVID-19 research. We evaluated its ability to model the direct and indirect effects...
168. SCALE: State-Calibrated Latent Embeddings for JEPA Planning in the Right Geometry ​
Author: Jiaming Hu, Yan Zheng, Tian Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16287v1 Announce Type: new Abstract: Joint-embedding predictive world models plan by scoring predicted terminal embeddings against a goal embedding using a cost defined on the representation itself. Two prominent strategies for obtaining non-collapsed representations are to inherit a pret...
169. Advancing Open and Reproducible Relational Learning: RelArena-$\alpha$, TabPFN-Rel and RPI ​
Author: Adrian Hayler, Klemens Fl"oge, Alan Arazi, Rishabh Ranjan, Jure Leskovec, Felix Birkel, Brendan Roof, Anurag Garg, Kristina Collins, Lydia Sidhoum, Jonas K"ubler, Siyuan Guo, Oscar Key, Jan Hendrik Metzen, Rylee Grace, David Salinas, Arthur Cahu, Simon Bing, Benjamin J"ager, Tuana \c{C}elik, Mihir Manium, Vitor Monteiro, Jake Robertson, Jerry Chen, Eliott Kalfon, Tom'as Pereda, Lilly Wehrhahn, Dominik Safaric, Tobias Schroeder, Georg Grab, Diana Kriuchkova, Clara Cornu, Philipp Singer, Nick Erickson, Vahid Balazadeh, Marie Salmon, Simone Alessi, K"ur\c{s}at Kaya, Philipp Jund, L'eo Grinsztajn, Yann LeCun, Bernhard Sch"olkopf, Madelon Hulsebos, Lennart Purucker, Sauraj Gambhir, Frank Hutter, Noah Hollmann
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16319v1 Announce Type: new Abstract: This first release of Prior Labs in relational learning shows our continued commitment to open science. We open-source three pieces of software that we expect to accelerate research in the field towards meaningful real-world impact. We aim to steer fur...
170. Transfer Learning of Keystroke Dynamics for Cross-Device User Authentication ​
Author: Nuwan Kaluarachchi, Sevvandi Kandanaarachchi, Kristen Moore, Arathi Arakala, Conrad Sanderson
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.HC
arXiv:2608.16334v1 Announce Type: new Abstract: Keystroke dynamics (typing patterns) can be used as a behavioural biometric modality for user authentication, with applications such as fraud prevention. While the modality has been shown to work well for single device authentication, its application t...
171. Task-Anchored Representation Shaping for Pre-Trained Model-Based Continual Learning ​
Author: Zhiming Xu, Huiyu Yi, Zhen-Hao Xie, Baile Xu, Furao Shen, Jian Zhao, Suorong Yang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16345v1 Announce Type: new Abstract: Pre-trained models (PTMs) provide a strong foundation for continual learning by offering stable representations that facilitate lightweight adaptation to new tasks. However, adapting well to each task does not ensure reliable inference over all learned...
172. OceanDepths: A Global Dataset of Paired Subsurface and Surface Ocean Observations ​
Author: Simon Donike, Ruben Cartuyvels, Antonino Ian Ferola, Elisa Carli, Diego Fernandez Prieto, Marie-Helene Rio
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.16373v2 Announce Type: new Abstract: Despite comprising over 70% of its surface, the world's oceans are critically underobserved compared to the land surface or the atmosphere. Understanding the global ocean requires jointly observing its surface and subsurface structure, yet no standardi...
173. Coverage-Maximizing Multinomial Subset Routing under Operational Constraints ​
Author: Quan Zhou, Yiyan Huang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16375v1 Announce Type: new Abstract: We introduce Multinomial Subset Routing (MSR), a new online routing framework over $K$ experts in which the learner keeps a multinomial routing policy instead of a deterministic subset of experts. At each round, the learner samples $M$ experts i.i.d. f...
174. FETERS: Few-Shot Early Time-Series Classification via Effective Ratio Selection ​
Author: Chen-An Tai, Yujia Wu, Vincent S. Tseng
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16385v1 Announce Type: new Abstract: Early time-series classification (ETSC) aims to make accurate predictions from partially observed time series as early as possible. Although various stopping mechanisms and feature learning strategies have been developed for ETSC, most existing methods...
175. SoftModel: A Neural Model That Grows Its Own Topology -- Governed Structural Growth for Continual In-Service Learning ​
Author: Zhoumin Xie
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16409v1 Announce Type: new Abstract: Today, a neural system is almost always used in two phases -- trained, then deployed -- and in that regime it freezes twice: training ends, and the topology itself was never a degree of freedom. We take the opposite premise as an axiom -- total plastic...
176. TRACE-CASH: Trial-History-Conditioned Reinforcement Learning for Adaptive Configuration Exploration in Time-Series CASH ​
Author: Yu-Han Huang, Yujia Wu, Vincent S. Tseng
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16410v1 Announce Type: new Abstract: Combined algorithm selection and hyperparameter optimization (CASH) searches a conditional space in which the selected model determines which hyperparameters are active. In time-series forecasting, temporal choices, chronological validation, and costly...
177. Evolving Executable Pipeline Programs for AutoML with Language Models ​
Author: Sofoklis Kitharidis, Cor J. Veenman, Jan N. van Rijn, Thomas B"ack, Niki van Stein
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2608.16416v1 Announce Type: new Abstract: Automated machine learning (AutoML) systems search for pipelines within a space of preprocessing operators, learners, and hyper-parameters specified in advance: they can select and tune known components, but cannot produce structure outside that space....
178. PertMind: Eliciting Emergent Biological Reasoning in LLM via Reinforcement Learning on Cellular Perturbation Data ​
Author: Zhenchao Tang, Xiaogang Xu, Tianxu Lv, Jiahui Guan, Jiale Zhou, Haohuai He, Zhi Song, Hanbo Huang, Jiehui Huang, Jiafei Wu, Zhe Liu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.QM
arXiv:2608.16419v1 Announce Type: new Abstract: Large language models can describe mechanisms, yet scalable post-training still depends on costly, manually curated biological reasoning traces. Here we show that cellular perturbation atlases can instead become reinforcement-learning environments, whe...
179. Localized TabICLv2: Scaling Tabular In-Context Learning through k-NN ​
Author: Beimnet Bekele Guta
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16429v1 Announce Type: new Abstract: Foundational models for tabular data have made significant progress in recent years, with TabICLv2 reporting state-of-the-art performance on several tabular classification tasks. However, full-context tabular ICL still suffers from attention cost that ...
180. Reference-free logged energy-oracle recovery for neural approximations of symmetric coercive variational problems: conforming Riesz reconstruction and archive-level selection ​
Author: Karim Bounja, Lahcen Laayouni, Boujemaa Achchab, Abdeljalil Sakat
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.16473v1 Announce Type: new Abstract: Neural PDE training yields a finite checkpoint archive, yet its logged energy errors are inaccessible without the exact solution, while loss-based selection does not necessarily recover the logged energy oracle. For admissible neural approximations of ...
181. Pallas: A Proactive KV Cache Migration Framework for LLM Inference in AI-RAN ​
Author: Tianhang Ding, Jianchun Liu, Hongli Xu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16477v1 Announce Type: new Abstract: AI-RAN brings large language model (LLM) serving close to mobile users, but cellular handover can separate an active request from its inference state: the user attaches to a target base station (gNB) while the large and growing key-value (KV) cache rem...
182. Graph Machine Learning: An Opportunity for Power Systems ​
Author: Martin Sadric, Sebastian P"utz, Christian Nauck, Veit Hagenmeyer, Frank Hellmann, Dirk Witthaut, Benjamin Sch"afer
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE, cs.SY, eess.SY
arXiv:2608.16494v1 Announce Type: new Abstract: Modern power systems face growing operational complexity driven by the integration of renewable energy sources, decentralization, and the need for real-time decision-making across a wide range of timescales. Addressing these challenges traditionally re...
183. When Tool-Backed Skill Retrieval Fails: Source-Style Collapse in Executable Capability Retrieval ​
Author: Yiqi Liu, Joseph James, Yang Wang, Chenghao Xiao, Chenghua Lin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2608.16502v1 Announce Type: new Abstract: Large-scale agents increasingly rely on retrieval to access external capabilities. We study this retrieval gate in structured tools and APIs, a measurable class of tool-backed executable skills that must be surfaced before an agent can plan, incorporat...
184. One Residual with Three Reuses: A Wristband Front End for Gesture Sensing ​
Author: Sam Rifaki
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.HC
arXiv:2608.16542v1 Announce Type: new Abstract: Continuous wrist-worn hand sensing for gesture interfaces and motor symptom monitoring needs an always-on front end that fits inside a coin-cell power budget while pairing a micro-electro-mechanical-systems (MEMS) inertial measurement unit (IMU) with a...
185. Learning Generalizable Reconstruction of High-Dimensional Neural Dynamics ​
Author: Anima Kujur, Zahra Monfared
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16569v1 Announce Type: new Abstract: Accurate reconstruction of long-duration neural recordings is challenging because local field potentials (LFPs) are high-resolution, multichannel, transient, and variable across subjects. We present PCA-DMD, a scalable operator-theoretic framework that...
186. Variational Outlier-Robust Gaussian Process Regression with Generative Modeling ​
Author: Arslan Majal, Aamir Hussain Chughtai
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16606v1 Announce Type: new Abstract: Outliers can substantially distort Gaussian process regression (GPR) due to its conventional Gaussian observation likelihood, leading to inaccurate model learning and prediction. To address this limitation, this article introduces a generative GPR mode...
187. Hoeffding adaptive splitting trees for data stream classification with concept drift and ensemble learning ​
Author: Daniel Nowak Assis, Jean Paul Barddal, Fabr'icio Enembreck
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16659v1 Announce Type: new Abstract: Ensembles of decision trees are well-established methods for data stream classification. In ensemble learning, Hoeffding Trees are widely adopted as base learners, performing periodic split attempts according to the Hoeffding bound. Recent studies, how...
188. UniTAC: Universal Task-Aware Compression via Weighted Distortion Measures ​
Author: Homa Esfahanizadeh, Matin Mortaheb, Jinfeng Du, Harish Viswanathan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, cs.MM, math.IT
arXiv:2608.16696v1 Announce Type: new Abstract: Physical AI systems such as autonomous vehicles and robots rely on timely exchange of high-dimensional sensory signals under tight bandwidth, latency, and energy budgets. Because the task driving downstream decisions evolves over time, a task-specific ...
189. Learning to Unlearn: Machine Unlearning via Learning the Unlearning Behaviors ​
Author: Hang Zhang, Kaifeng Zhang, Yixiao Ma, Weijie Xu, Ye Zhu, Kai Ming Ting
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16700v1 Announce Type: new Abstract: Various machine unlearning techniques have been developed in response to privacy legislation requirements, enabling individuals to exercise their legal right to have their data $D_f$ removed from a machine learning model. This process is typically acco...
190. The Ethical Decision Head: Operationalizing Normative Ethics in Autonomous Vehicles via Reinforcement Learning from Human Feedback ​
Author: Thomas Mbrice, Ammar Ali, Sami Mian, Khai Hern Low, Eric Chen, Arshia Aghajani, Wolf Sch"afer, Amin Shirangi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16710v1 Announce Type: new Abstract: As autonomous vehicles (AVs) approach Level 4 and Level 5 operational capability [SAE International, 2018], their on- board decision systems must handle not only safety-critical locomotion but also their subsequent moral weight. This paper details the ...
191. Le Critique: Privileged Value Functions for LLM Reinforcement Learning ​
Author: Siddarth Venkatraman, Matthieu Dinot, Laurence Aitchison
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16739v1 Announce Type: new Abstract: Reinforcement learning algorithms for Large Language Models (LLMs) are largely distinguished by their variance reduction strategy. Group-relative methods like GRPO reduce gradient variance by sampling multiple rollouts per prompt, but provide only sequ...
192. Would this change your answer? Evaluating Explanations of LLM Behavior In The Wild with Counterfactual Experiments ​
Author: Adam Karvonen, Euan Ong, Subhash Kantamneni, Samuel Marks
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16747v1 Announce Type: new Abstract: Many areas of AI research, such as language model interpretability and chain of thought faithfulness, seek to explain model behaviors. But what constitutes a "good" explanation? In this work, we evaluate explanations through the lens of counterfactual ...
193. On the Principles Behind Neural Network Optimizers ​
Author: Yushun Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2608.16760v1 Announce Type: new Abstract: Reliable optimization is central to neural network (NN) training, yet Adam, the default optimizer for modern LLMs, rests on a fragile foundation. This thesis develops a principled grounding for Adam and motivates new designs. First, we revisit Adam's d...
194. Beyond $L_2$: Generalizing Abductive Latent Explanations to Diverse Prototype-Based Architectures ​
Author: Jules Soria, Alban Grastien, Romain Xu-Darme, Julien Girard-Satabin, Zakaria Chihani, Daniela Cancila
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16773v1 Announce Type: new Abstract: Prototype-based neural networks are hailed as interpretable-by-design architectures. Recently, Abductive Latent Explanations (ALE) were introduced to provide formal, mathematically guaranteed explanations that leverage the intrinsic structure of these ...
195. GEO-Flag: Detecting and Measuring GEO-Optimized Web Content ​
Author: Junjie Chu, Ye Leng, Mingjie Li, Yun Shen, Xinyue Shen, Yang Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.IR
arXiv:2608.16824v1 Announce Type: new Abstract: Generative Engine Optimization (GEO) modifies web content to increase its likelihood of being selected and cited by generative search engines. This can give strategically optimized pages visibility disproportionate to their authority or relevance and e...
196. CaliBench: Are the Stochastic Dynamics of Video World Models Physically Calibrated? ​
Author: Jonathan Sadeghi, Jenny Seidenschwarz, Jesse Allardice, Sirish Srinivasan, Benjamin Graham, Jeffrey Hawke
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.16829v1 Announce Type: new Abstract: Video world models approximate the stochastic distribution of physical outcomes through generative sampling, but existing benchmarks score individual generations or compare distributions coarsely over a whole dataset, leaving the fine-grained aleatoric...
197. Time-Aware Validation of Machine Learning Fuel Consumption Models: Evidence from 1\,Hz Operational Data, CCGS \textit ​
Author: Samarasimha Reddy Chittamuru, Ayhan Akinturk, Allison Kennedy, Joshua Barnes, Matthew Hamilton
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16833v1 Announce Type: new Abstract: Ship fuel consumption (SFC) prediction supports vessel operation optimisation, emissions estimation, and decision support systems (DSS) for sustainable maritime transportation. Numerous data-driven fuel models have been developed over the past two deca...
198. Proteus: Incremental Memory Activation for Long-Context Sequence Modeling ​
Author: Reza Bayat, Ali Behrouz, Vahab Mirrokni, Aaron Courville
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.16844v1 Announce Type: new Abstract: The quadratic cost of attention-based sequence models for long contexts has motivated a growing line of research on memory-based models that can compress context into a compact state. However, most existing memory models expose a static memory througho...
199. Data-Efficient and Interpretable Classification of Circulating Tumor Cell Phenotypes in Microfluidic Devices via Deep Learning ​
Author: Serena Su, Yifan Wang, Senwei Liang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16870v1 Announce Type: new Abstract: Accurate classification of circulating tumor cell (CTC) phenotypes can provide valuable information for assessing metastatic potential. Label free microfluidic devices provide a hydrodynamic obstacle course that transforms subtle biophysical characteri...
200. A Data-Efficient Analytical Prior Machine Learning Framework for Sound Reduction Frequency Prediction in Helmholtz Resonators ​
Author: Jiaming Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16873v2 Announce Type: new Abstract: High-fidelity finite-element simulations can provide accurate numerical predictions for side-branch resonators, but large simulation datasets are expensive to generate and purely data-driven surrogates may become unreliable when simulation-labelled dat...
201. Q-based Variational Inverse Reinforcement Learning ​
Author: Ondrej Bajgar, Peter Tisnikar, Alessandro Abate, Konstantinos Gatsis, Maike Osborne
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.16888v1 Announce Type: new Abstract: The development of safe and beneficial AI requires that systems can learn and act in accordance with human preferences. However, explicitly specifying these preferences by hand is often infeasible. Inverse reinforcement learning (IRL) addresses this ch...
202. Auxiliary uncertainty signals for LLM-assisted systematic review screening: a benchmark across eight Cohen drug-class reviews ​
Author: Arya Rahgozar, Pouria Mortezaagha
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.DL, cs.IR, cs.LG
arXiv:2608.14551v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for title-abstract screening in systematic reviews, but their decisions lack calibrated uncertainty. We show that an auxiliary BERT+GCN classifier supplies a structured uncertainty signal that improv...
203. When to Communicate: Belief Distributions and KL Divergence for Principled Gating in Multi-Agent RL ​
Author: Teoman Kaman
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.14559v1 Announce Type: cross Abstract: Effective communication in multi-agent reinforcement learning requires agents to decide not only \textit{what} to communicate, but when? Existing approaches either communicate at every timestep or learn a binary gate through REINFORCE policy gradient...
204. Longitudinal and Graph-Augmented Prediction of Adolescent Substance Use Onset in the ABCD Study ​
Author: Yixuan He, Jinni Su, Yun Kang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG, stat.AP
arXiv:2608.14578v1 Announce Type: cross Abstract: Early identification of adolescent substance-use risk is an important prevention challenge, yet the relative value of baseline characteristics, longitudinal trajectories, and relational context remains unclear. Using data from approximately 11,860 pa...
205. Characterization of Thermal Systems from Noisy and Low-resolution Measurements Using Dynamic Mode Decomposition ​
Author: M. E. P. Silva, L. S. Araujo, F. T. Colombo, A. Cunha Jr, S. da Silva
Published: 8/19/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG, eess.SP, physics.class-ph
arXiv:2608.14581v1 Announce Type: cross Abstract: Thermal monitoring in practical applications is often constrained by sparse sensing, measurement noise, and limited spatial resolution, which hinder the identification of heat transfer dynamics. In such settings, calibrating high-fidelity physical mo...
206. Evaluating the impact of adversarial traffic patterns on vanet communication using veins simulation ​
Author: Henry Agyapong
Published: 8/19/2026, 4:00:00 AM
Categories: cs.NI, cs.LG
arXiv:2608.14583v1 Announce Type: cross Abstract: Vehicular Ad Hoc Networks (VANETs) are a key component of intelligent transportation systems, enabling real-time communication between vehicles. However, their open and dynamic nature makes them highly vulnerable to adversarial behaviors that can dis...
207. 6G Native AI and Channel Foundation Models ​
Author: Shugong Xu, Jun Jiang, Yuan Gao
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.IT, cs.LG, math.IT
arXiv:2608.14591v1 Announce Type: cross Abstract: The integration of artificial intelligence (AI) and wireless communications is widely regarded as a core objective of sixth-generation (6G) systems. However, both the meaning of native AI and the type of AI capability that should be embedded into fut...
208. Demo: Real-time Generative Multicasting with On-Device Intent-aware Semantic Decomposition ​
Author: Xinkai Liu, Mahdi Boloursaz Mashhadi, Yi Ma, Rahim Tafazolli
Published: 8/19/2026, 4:00:00 AM
Categories: cs.NI, cs.LG, cs.MM
arXiv:2608.14600v1 Announce Type: cross Abstract: We present a demonstration for generative multicasting with on-device, intent-aware semantic decomposition. At the transmitter, DNN-based segmentation extracts a semantic map from the source video, decomposing it into multiple sub-signal classes base...
209. LLM Safety Alignment in Low-Resource Languages: A Systematic Literature Review ​
Author: Valdini Douglace Lemofouet, Blessing Ngozi Uzor, Paula Chikaodinaka Anyanwu, Danielle Blanche Kapsa, Sukairaj Hafiz Imam, P Sam Sahil, Abigail Oppong, Tassallah Abdullahi, Clemencia Siro, Idris Abdulmumin, Seid Muhie Yimam, Shamsuddeen Hassan Muhammad
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.14626v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved substantial progress in safety alignment, yet their safety guarantees remain significantly weaker in low-resource and multilingual settings than in high-resource languages. In this paper, we conduct a System...
210. Wolff-Parkinson-White Detection at 471:1 Class Imbalance: A Leakage-Controlled Study of the Data Bottleneck ​
Author: Nathael Altman
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.14633v1 Announce Type: cross Abstract: Wolff-Parkinson-White (WPW) syndrome is a congenital cardiac pre-excitation, clinically important and often missed on the resting 12-lead ECG. Detection is hard: the signature is subtle and the condition rare. We pool two public 12-lead corpora, PTB-...
211. Belayer: Efficient Fault Tolerance for LLM Agentic RL Training ​
Author: Jiecheng Zhou, Qinghao Hu, Peng Sun, Xingcheng Zhang, Weiming Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2608.14635v2 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly trained with reinforcement learning in long-horizon, sandboxed environments. Unlike conventional RL, agentic RL couples GPU-intensive rollout engines with stateful environment containers whose action...
212. Stop Indexing at Full Precision: Revisiting Clustering for Vector Embeddings ​
Author: Leonardo Kuffo, Peter Boncz
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DB, cs.AI, cs.LG
arXiv:2608.14648v1 Announce Type: cross Abstract: In this study, we revisit three widely used techniques in vector search and utilize them to optimize vector embedding indexing through clustering: dimensionality reduction, quantization, and dimension pruning. We propose an indexing pipeline in which...
213. When Uncertainty Isn't Enough: An Empirical Study of Self-Correction in Code Generation ​
Author: Pranav Rakasi, Maanas Lalwani, Arnav Srivastava, Arya Palanivel, Tinuade Adeleke, Ruizhe Li, Sean Wu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE
arXiv:2608.14659v1 Announce Type: cross Abstract: Large language models for code generation often produce incorrect solutions without reliable indicators of failure. We study whether uncertainty estimation methods developed for natural language transfer to code generation, and whether such signals c...
214. Does the Heart Show Your Pain? Tackling the X-ITE Pain Challenge with Self-Supervised ECG Representation Learning ​
Author: Dominika Kunc, Przemys{\l}aw Kazienko, Stanis{\l}aw Saganowski
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG
arXiv:2608.14662v1 Announce Type: cross Abstract: Accurate recognition of pain using physiological signals remains a challenging problem due to pain's subjective nature and high inter-individual variability. In this study, we investigate self-supervised representation learning (SSL) methods applied ...
215. Phase-Aware CNN for Real-Time 5G/6G Channel Estimation with Hardware-in-the-loop Validation ​
Author: Javad Zolfaghari-Bengar, Rakibul Rony, Elisa Gomez-de-Lope, Alejandro Villena-Rodriguez, Abhinav Mahadevan, Nicolas Kourtellis
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.14676v1 Announce Type: cross Abstract: In 5G/6G wireless systems, accurate and timely channel estimation is critical to ensure reliable communication under complex, fast-changing radio conditions. This work focuses on pilot-based channel estimation using deep learning to reconstruct both ...
216. Offline Ambient-Controlled Latent Diffusion: Architecture, Telemetry, and On-Device Evaluation ​
Author: Lech Kalinowski, Artur Morys-Magiera, Piotr Mi{\l}kowski
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG
arXiv:2608.14677v1 Announce Type: cross Abstract: Most mobile image-generation applications are thin clients over cloud services, leaving outputs hard to audit. We present an Android latent-diffusion application that runs entirely on-device and is driven by the ambient-light sensor rather than a tex...
217. A Low-Cost IoT Device for Environmental Monitoring and Embedded Solar Forecasting with On-Device Incremental Learning ​
Author: Erick Michel Lara Pinal, Abhinav Das, Stephan Schl"uter
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.14698v1 Announce Type: cross Abstract: Hyperlocal meteorological sensing is essential for accurate solar photovoltaic forecasting, yet professional-grade meteorological stations require investments easily exceeding 1000~USD per node, making distributed deployments economically inaccessibl...
218. Deep Analog: Open-Set Film Emulation with Reference-Conditioned 3D LUTs ​
Author: Yitong Mu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.LG, cs.MM
arXiv:2608.14702v1 Announce Type: cross Abstract: Film emulation reproduces the look of an analog film stock on a new digital photograph. We target its open-set form -- matching any reference film frame from a single example -- with a 3D lookup table (LUT) predicted from that reference. Real-time im...
219. On Cross-Validation for Hyperparameter Optimization of Deep Learning Image Classifiers ​
Author: Ljubomir Buturovic (East Palo Alto, United States)
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.14705v1 Announce Type: cross Abstract: Hyperparameter optimization (HPO) can materially affect the performance of deep learning (DL) image classifiers, but there is little empirical guidance on how to derive the validation signal that drives it, especially for the small sample sizes commo...
220. Equilibrium Forcing: Adaptive Video Generation Without Noise Conditioning ​
Author: Hansen Jin Lillemark, Alex Rojas, Zachary Novack, Runqian Wang, Yilun Du, Yian Ma, Taylor Berg-Kirkpatrick, Rose Yu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.14706v1 Announce Type: cross Abstract: Standard autoregressive video generation algorithms based on Diffusion and Flow Matching rely on rigid training objectives and static sampling schedules, limiting inference procedures from adapting to the data. We introduce Equilibrium Forcing (EqF),...
221. Hardware-in-the-Loop Phase-Aware CNN for Real-Time 5G Channel Estimation ​
Author: Javad Zolfaghari-Bengar, Rakibul Rony, Elisa Gomez-de-Lope, Alejandro Villena-Rodriguez, Abhinav Mahadevan, Nicolas Kourtellis
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.CV, cs.IT, cs.LG, math.IT
arXiv:2608.14709v1 Announce Type: cross Abstract: This demo presents real-time AI-based uplink channel-estimation inference using data collected from a hardware-in-the-loop 5G platform. The data-collection setup integrates commercial RF signal generation, programmable channel emulation, an O-RAN Rad...
222. Which Question Is Your Attention Metric Answering? Attention Rows as Compositional Data ​
Author: Marios Papamichalis, Regina Ruane
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, math.ST, stat.TH
arXiv:2608.14712v1 Announce Type: cross Abstract: Each row of a transformer's attention matrix is a probability distribution over tokens, and in trained models most of that probability lands on a single \emph{sink} token, usually the first. Standard tools for comparing attention rows (cosine similar...
223. Koopman early warning signals for bifurcation and rate-induced tipping ​
Author: Juan Nathaniel, Carla Roesch, Derek DeSantis, Parvathi Kooloth, Hang Fan, Valerio Lucarini, Anastasia Romanou, Pierre Gentine
Published: 8/19/2026, 4:00:00 AM
Categories: nlin.CD, cs.LG
arXiv:2608.14716v1 Announce Type: cross Abstract: Abrupt transitions in complex systems are often preceded by early warning signals. However, most indicators rely on the notion of critical slowing down and do not generally extend to rate-induced tipping where transitions can occur without local loss...
224. Local Gains and Fixed-Assignment Set Losses in Shared Set Decoders ​
Author: Ze Zhang, Yang Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.14717v1 Announce Type: cross Abstract: A query-relation deletion can improve the edited slot while reducing the utility of the prediction set that contains it. We study this tension in two related ResNet-50 DETR-family checkpoints using recorded, selection-conditional evidence from 710 pa...
225. A Vision Transformer for ECG-Based Detection of Left Ventricular Systolic Dysfunction Across Multiple Clinical Sites ​
Author: Burcu Ozek, Aruna Mohan, David Vorchheimer, Daniel Weiss, Eyal Kedar, Tamar Sobol, Or Zilbershot, Fatemeh Afghah
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.14723v1 Announce Type: cross Abstract: Reduced left ventricular ejection fraction (LVEF) is frequently asymptomatic and often detected only after advanced heart failure develops. Electrocardiograms are recorded routinely yet underused for this condition, because reduced LVEF has no single...
226. Privacy-Preserving Dataset Curation for Kuala Lumpur Urban Traffic: Grounded Vision-Language Detection with Spatial Vehicle-Context Filtering ​
Author: Mohammed Abdul Al Arafat Tanzin, Rudzidatul Akmam Dziyauddin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.14724v1 Announce Type: cross Abstract: The rapid advancement of intelligent transportation systems and autonomous driving relies heavily on multi-modal urban traffic datasets. However, curating high-fidelity video imagery in complex tropical urban environments---specifically Kuala Lumpur,...
227. IP Protection in the Era of Visual Generative AI: A Survey ​
Author: Zhuan Shi, Shunchang Liu, Alireza Dehghanpour Farashah, Qian Yang, Han Yu, Cao Yang, Chaochao Chen, Yuping Yan, Yaochu Jin, Golnoosh Farnadi, Lingjuan Lyu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.CR, cs.LG
arXiv:2608.14730v1 Announce Type: cross Abstract: The rapid evolution of visual generative AI has introduced a wide range of intellectual property risks, spanning the unauthorized learning, reproduction, extraction, misuse, and redistribution of protected data and model assets. To address these risk...
228. PandasCorpus: A Resource of Real-World Pandas Workflows and Usage Patterns ​
Author: Syrym Abdikhan, Mazhar Hameed
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SE, cs.LG
arXiv:2608.14742v1 Announce Type: cross Abstract: Pandas has emerged as the de facto library for data processing and machine learning, widely used for tasks, such as data loading, transformation, and analysis. Despite its ubiquity, there has been limited systematic investigation into how Pandas is u...
229. The Note-Chord-Voice Framework: Structured Source Separation and Causal Inference for EV Charging Data ​
Author: Jiajie Chen, Jinfeng Li
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.LG, cs.SD
arXiv:2608.14756v1 Announce Type: cross Abstract: Real-world EV charging data exhibit three interlocking pathologies: hardware fragmentation (network timeouts and billing resets split sessions), physical violations (independent energy/duration models produce impossible states like 50 kWh in 10 min o...
230. CFR without Unbiasedness: Deterministic Guarantees for Persistent Public-Chance Schedules ​
Author: Jiaxing Guo, Lei Ye
Published: 8/19/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2608.14761v1 Announce Type: cross Abstract: At a finite public-chance cut, counterfactual regret minimization (CFR) must choose how many outcomes to evaluate before each regret update. Exact evaluation processes the full cut at one strategy profile; persistent partial evaluation processes a fi...
231. Cross-Modal Ultrasound-MRI Learning for Fetal Brain Ventricular Volumetry and Abnormality Screening ​
Author: Yuhao Huang, Yuanji Zhang, Yuhuan Lu, Dong Ni, P. Ellen Grant, Davood Karimi
Published: 8/19/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG
arXiv:2608.14763v1 Announce Type: cross Abstract: Assessment of ventriculomegaly (VM) on fetal brain ultrasound relies primarily on measuring lateral ventricular atrial width on standard planes, which is operator-dependent and may not fully reflect the overall ventricular enlargement. Fetal brain MR...
232. Beyond Boundary Noise: Aggregated Aleatoric Uncertainty Fails to Capture Presence Ambiguity in 3D Lung Nodule Segmentation ​
Author: Simon Baur, Arne Schernich, Ekin B"oke, Wojciech Samek, Jackie Ma
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.14766v1 Announce Type: cross Abstract: Uncertainty estimation is critical for the safe clinical deployment of deep learning in medical image segmentation, with aleatoric uncertainty theoretically designed to capture irreducible data ambiguity. However, whether entropy-based measures refle...
233. Uncertainty Identifies Difficult Samples Across Methods: A Multi-Task Study on a Heterogeneous Skin Lesion Dataset ​
Author: Leon Koole, Jiapan Guo, Matias Valdenegro-Toro
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.14768v1 Announce Type: cross Abstract: Skin lesion classifiers can be confidently wrong on the cases that matter most, so knowing when a prediction should not be trusted is clinically as useful as the prediction. We study uncertainty quantification on a dataset pooled from many ISIC sourc...
234. From Errors to Proofs: Minimal-Core-Guided Repair for Neuro-Symbolic Constraint Solving ​
Author: Dipankar Sarkar
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.LO, cs.PL, cs.SC, math.OC
arXiv:2608.14771v1 Announce Type: cross Abstract: Making language models solve constraint problems reliably often means having them translate the problem into a formal specification and delegating the search to a sound solver. But the translation is itself a language-model task, and an unfaithful tr...
235. AMPLIFAI: A Multiphase CT Dataset for Benchmarking Clinical Reasoning in LI-RADS Assessment of Liver Lesions ​
Author: Pranav Kulkarni, Nikhil Shah, Amritansh Suryavanshi, Jana Delfino, James Tonascia, Jade Wong-You-Cheong, Barton Lane, Joseph Chirico, Jeffrey D. Hirsch, Ang Li, Heng Huang, Florence X. Doo
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.14778v1 Announce Type: cross Abstract: Hepatocellular carcinoma (HCC) is the third leading cause of cancer-related mortality worldwide, with early detection improving survival from 70%. The standardized LI-RADS criteria establish a biopsy-free, fully imaging-based framework that can serv...
236. Prompting is not enough: supervised baselines and leakage control for measuring shared decision-making with LLMs in pediatric encounters ​
Author: Bernardo Modenesi, Jody Lin, Kimberly Kaphingst, Angela Zhu, Maya Wheeler, Peilu Zhang, Angela Fagerlin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.14792v1 Announce Type: cross Abstract: Objectives: To determine whether zero-shot prompting of a large language model (LLM) is sufficient to detect shared decision-making (SDM) behaviors in real clinical encounters, and whether supervised learning adds value under patient-grouped, nested ...
237. Zero-Shot Adaptation of Medical Vision Foundation Models for High-Frequency Micro-Ultrasound Prostate Segmentation ​
Author: Ayusha Abbas, Saram Abbas, Kabita Adhikari
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.14796v1 Announce Type: cross Abstract: Prostate cancer claims a life every 80 seconds. Early detection is needed to prevent disease progression, and both PSA density calculation and biopsy decisions rely on knowing the exact boundary of the gland. Conventional ultrasound at 6-12 MHz blurs...
238. Do LLMs Know What to Ask and When? Evaluating Multi-Turn Information Seeking ​
Author: Yepeng Huang, Jiawen Zhang, Michelle Dai, Xiaorui Su, Shanghua Gao, Zi Wang, Marinka Zitnik
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.14808v1 Announce Type: cross Abstract: When a user question is underspecified, a capable model should recognize that its context is insufficient, identify the missing information, ask for it, and respond only once that information determines a unique answer. We formalize multi-turn inform...
239. What Makes a Good Layer? Assessing the Layer-Wise Intrinsic Properties of Music Foundation Models ​
Author: Angelos-Nikolaos Kanatas, Yuexuan Kong, Pablo Alonso-Jim'enez, Xavier Serra, Dmitry Bogdanov
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS
arXiv:2608.14819v1 Announce Type: cross Abstract: Music foundation models are commonly used as frozen audio feature extractors, yet selecting which layer to extract from remains largely heuristic. Current practice defaults to fixed depths or multi-layer fusion, with limited understanding of why cert...
240. A Parameter-Free Few-Shot Evaluation for Elephant Vocalisation Classification ​
Author: Christiaan M. Geldenhuys, Thomas R. Niesler
Published: 8/19/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, cs.SD, q-bio.QM
arXiv:2608.14824v1 Announce Type: cross Abstract: We present a parameter-free episodic evaluation of nearest-centroid classification for elephant vocalisations on fixed pretrained acoustic embeddings, across the Elephant Voices (EV) and Linguistic Data Consortium (LDC) datasets. Rather than asking w...
241. Explainability Boosted Anomaly Detection Framework for O-RAN based NextG Networks ​
Author: Nurullah Aksu, Ali Fuat Sahin, Semiha Tedik Ba\c{s}aran
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.CR, cs.LG
arXiv:2608.14826v1 Announce Type: cross Abstract: The wireless networks have historically faced significant security vulnerabilities, necessitating advanced anomaly detection mechanisms, especially as networks evolve towards 6G and beyond. This study introduces an advanced anomaly detection framewor...
242. MINT: Min-Selection Preference Distillation for Balanced Multi-Objective Alignment ​
Author: Tony Tu, Sayan Chakraborty, Ruomeng Xu, Tony Qin, Austin Tian
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.14828v1 Announce Type: cross Abstract: Aligning a language agent to several objectives at once is a persistent failure mode of preference-based training: when objectives are combined additively, optimization collapses onto whichever is cheapest to improve and sacrifices the rest, so a sup...
243. Discovering High-Quality Chess Puzzles with Offline Reinforcement Learning ​
Author: Allen Nie, Anirudhan Badrinath, Nicholas Tomlin, Timothy Dai, Carissa Yip, Rose E Wang, Emma Brunskill, Chris Piech
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.14851v1 Announce Type: cross Abstract: Learning and skill mastery require extensive and deliberate practice. In many learning settings, producing high-quality pedagogical materials can require a high level of domain expertise and be very time-consuming. Pedagogical materials often need to...
244. BRAID: Learning Equilibrium Maps in Interdependent Security Games via Weight-Tied Iterative Graph Neural Networks ​
Author: Elnaz Nowrouzi, Zhiqun Zuo, Xueru Zhang, Mohammad Mahdi Khalili
Published: 8/19/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2608.14856v1 Announce Type: cross Abstract: Computing Nash equilibria in interdependent security (IDS) games on networks is computationally expensive: best-response dynamics may need hundreds of iterations per instance, and downstream tasks such as auditing, stress-testing, and incentive desig...
245. STAR-FL: Secure Federated Learning with Spatial-Temporal Analysis and Robust Aggregation ​
Author: Nawrin Tabassum, Yanzhao Wu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.14861v1 Announce Type: cross Abstract: Data poisoning attacks pose serious security threats to Federated Learning (FL) systems in Computer Vision. Despite growing research attention, two key challenges remain for existing defense techniques: (1) accurately distinguishing between benign an...
246. ARISE: An adaptive residual-informed stability ensemble for feature selection in small-sample biomedical omics ​
Author: Zardad Khan, Amjad Ali, Naz Gul, Sheema Gul, Saeed Aldahmani
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.14866v1 Announce Type: cross Abstract: Objective: Small-sample molecular classification requires feature selectors that identify predictive, stable, and nonredundant subsets for binary and multiclass outcomes. We propose ARISE (Adaptive Residual-Informed Stability Ensemble), which integra...
247. When do machine-learned exchange-correlation improvements inherit into density-functional tight binding? ​
Author: Can Polat, Mustafa Kurban, Erchin Serpedin, Hasan Kurban
Published: 8/19/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG, physics.chem-ph, quant-ph
arXiv:2608.14875v1 Announce Type: cross Abstract: Machine-learned exchange-correlation functionals correct band gaps at near-semilocal cost, while density-functional tight binding reaches the $10^3$-$10^6$-atom regime; combining them assumes that a better parent yields a better parameterization, but...
248. Workspace Topology as an Attack Vector in Agentic Coding Assistants ​
Author: Alexandre G. R. Day, Pradeep Yadlapalli, Sriram Venkatapathy, Thomas Paniagua, Nick Raines, Sahil Wadhwa, Himanshu Kumar, Andy Luo, Sudeep Panyam, Rikhiya Ghosh, Pranab Mohanty, Giri Iyengar
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG
arXiv:2608.14876v1 Announce Type: cross Abstract: Agentic coding assistants are finding widespread use, not just in new code development but in quickly ingesting and leveraging third-party code. This opens up a risk of malicious code being ingested as these coding tools operate with broad filesystem...
249. Interpretable Cross-Lingual Alignment in Small Language Models: Probing Cultural and Pragmatic Reasoning in Japanese-English Bilingual LLMs ​
Author: Florian Braun
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.14896v1 Announce Type: cross Abstract: Large language models work well on English and behave in poorly understood ways on languages typologically far from it. Japanese is a clean example, where evaluation still leans on translation quality and JGLUE-style benchmarks, which roll lexical, s...
250. Optimal Watermark Localization in Mixed-Source Large Language Model Texts ​
Author: Jose H. Blanchet, T. Tony Cai, Xiang Li, Hao Liu, Qi Long, Weijie J. Su
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ME, cs.CL, cs.LG, stat.ML
arXiv:2608.14906v1 Announce Type: cross Abstract: Watermarking provides a principled way to authenticate text generated by large language models (LLMs). In practice, however, the final text may be mixed-source, with watermark evidence surviving at only a subset of token positions after rewriting, in...
251. Distinguishing AI-Generated Music from Edited Audio as a Hard-Negative Robustness Task ​
Author: Alexandru-Stefan Morosanu, Valerian Cecan, Stefan-Daniel Achirei, Laura Erhan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG
arXiv:2608.14916v1 Announce Type: cross Abstract: AI-generated music detectors are commonly evaluated against original songs, but real-world uploads are often remixed, re-encoded, pitch-shifted, or otherwise edited. These edited versions form a difficult negative class: they are not generated by AI,...
252. LLMs Can Predict Failure Risk, But Struggle to Predict Which Collaboration Protocol Pays Off: Cost-Aware Protocol Routing Across Reasoning Tasks ​
Author: Chih-Hsuan Yang, Jingyan Jiang, Cheng-Hau Yang, Vikram Vasudevan, Huihuo Zheng, Venkatram Vishwanath, Rajeev Thakur
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.14927v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) systems can improve reasoning by spending more computation, but deployment requires deciding when extra collaboration is worth its cost. We isolate this decision by running every problem under four protocols whi...
253. Training Leaves Traces: Centered Residual Signatures for Language Model Lineage Verification ​
Author: Aman Singh Thakur, Rayan Khoury
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.14929v1 Announce Type: cross Abstract: Open-weight language models are fine-tuned, quantized, pruned, and merged, yet their provenance is often undocumented. We study data-free white-box lineage verification: can weights alone reveal whether two compatible model checkpoints share ancestry...
254. Developing an Offshore Machine Learning Surface Layer Scheme ​
Author: Susan Dettling, Sue Ellen Haupt, Thomas Brummet, Patrick Hawbecker, Branko Kosovi'c, David John Gagne
Published: 8/19/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG
arXiv:2608.14935v1 Announce Type: cross Abstract: Turbulent fluxes between the surface and the atmosphere are typically parameterized using empirically fit relationships. Here we test machine learning techniques for fitting the relationship for the offshore environment. To do that, data from three o...
255. SkillComposer: Learning Reusable Skills for Natural-Language Robot Programming ​
Author: John Woods, Hasti Seifi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.RO, cs.CL, cs.LG
arXiv:2608.14944v1 Announce Type: cross Abstract: Natural-language interfaces can lower the barrier to programming robots, but existing systems struggle when users request complex tasks. While large language models (LLMs) perform well with simple commands, they often struggle to generate code for mu...
256. T-LLM Compiler: Trusted LLM-based Code Optimization and Verification Framework ​
Author: Zahra Fazel, Sunanda Gamage, Shayan Shirahmad Gale Bagi, Amir H. Ashouri, Tomasz S. Czajkowski, Bryan Chan, Reza Azimi, Yaoqing Gao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.PF, cs.PL
arXiv:2608.14953v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have opened opportunities to apply high-level code transformations to the field of code optimization, and it has since emerged as one of the most fundamental tasks for LLMs to perform; however, at prese...
257. A Deep Learning Model for Spatially Clustered Data via Differentiable Cluster Assignment ​
Author: Kexuan Li, Weidong Ma
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.14968v1 Announce Type: cross Abstract: We consider nonparametric regression when the association between a response and its covariates changes across an unknown partition of a spatial domain. The proposed estimator learns the partition and the cluster-specific regression functions jointly...
258. Spinning Conformal Correlators from Neural Networks ​
Author: Manas Dogra, James Halverson, Joydeep Naskar
Published: 8/19/2026, 4:00:00 AM
Categories: hep-th, cs.LG
arXiv:2608.15001v1 Announce Type: cross Abstract: We construct spinning conformal fields from neural networks and the embedding formalism, computing their two-, three- and four-point functions in examples, building on scalar conformal field techniques introduced in \cite{Halverson:2024axc}. For a pa...
259. NPU Offloading of a Frozen Visual Encoder for Robot Policy Training ​
Author: Hyojun Yun, Seungjae Won, Hyungpil Moon
Published: 8/19/2026, 4:00:00 AM
Categories: cs.RO, cs.AR, cs.LG
arXiv:2608.15002v1 Announce Type: cross Abstract: When a robot policy is trained for a new task or dataset, its visual encoder can be frozen and only its action generation module trained, reducing training cost. Freezing removes the encoder's backward pass, but its forward pass must still run at eve...
260. DualMiT-Net: Local-Global Transformer-Convolutional Fusion for Breast Mass Segmentation in Mammographic Regions of Interest ​
Author: Alibek Kamiluly, Milana Muratova, Yash Patel, Fan Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.15019v1 Announce Type: cross Abstract: Breast mass segmentation is an important step in computer-aided mammography, but it remains difficult because masses can have low contrast, irregular shapes, and boundaries that blend with surrounding breast tissue. To address this problem, we presen...
261. Prototype-Rectified Iterative Self-supervised Manifold Denoising under Severe Acoustic Shift ​
Author: Ashish Anand Shukla, Rini Smita Thakur, Aryan Das, Vinod K. Kurmi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SD, cs.LG
arXiv:2608.15037v1 Announce Type: cross Abstract: Audio-Text Foundation Models (ATMs) fail catastrophically under severe acoustic noise, yet existing adaptation strategies either rely on gradient-based Test-Time Adaptation (TTA), which reinforces noise rather than signal, or on prompt tuning that re...
262. Uncovering Hidden Leptonic Correlations with Flow Matching and Autoencoders ​
Author: Haruto Kitagawa, Satsuki Nishimura, Hajime Otsuka
Published: 8/19/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-th
arXiv:2608.15042v1 Announce Type: cross Abstract: We perform a global search for values of the Yukawa matrices and Majorana masses in the Type-I seesaw mechanism. Using flow matching, which is a generative artificial intelligence (generative AI) method, we generate a broad set of solutions reproduci...
263. RecurrentGPT: Expressive Depth through Recurrent Modulation in Transformers ​
Author: Amr Hegazy, Amr Alanwar, Mostafa Elhoushi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.15062v2 Announce Type: cross Abstract: Scaling transformer language models creates an inherent tension between expressivity and memory efficiency. While unique weights across layers preserve functional specialization---from input-grounding to abstract refinement---they incur a substantial...
264. Distribution-free false-alarm calibration and chance-corrected spatial evaluation for industrial anomaly detection ​
Author: Jie Deng
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.15090v1 Announce Type: cross Abstract: Studies of industrial visual inspection commonly report the area under the receiver operating characteristic curve (AUROC) and the overlap between anomaly maps and defect masks. Neither measure specifies the false-alarm rate at a selected threshold, ...
265. Fast Test-Time Refinement for Robust Learned Image Compression ​
Author: Jiaming Liang, Chi-Man Pun, Weisi Lin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.15113v1 Announce Type: cross Abstract: Learned image compression (LIC) has demonstrated remarkable rate-distortion (RD) performance in benign settings. However, the high representational capacity endowed by deep neural networks (DNNs) comes at the expense of increased adversarial vulnerab...
266. Anatomy of a Quantized Agent: VRAM Stability and Forecasting in Code-Synthesis Agentic Workloads ​
Author: Anubhab Banerjee
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.DC, cs.LG
arXiv:2608.15117v1 Announce Type: cross Abstract: Analytical models of peak VRAM consumption for LLM inference decompose memory into weight-storage, KV-cache, and activation terms parameterized by step count, tool invocations, and context expansion. We evaluate this decomposition empirically within ...
267. Sufficient Dimesion Reduction via Generalized Stein's Lemma ​
Author: Ye Tian
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2608.15121v1 Announce Type: cross Abstract: Sufficient dimension reduction (SDR) seeks the minimal subspace of the predictors that captures the full conditional distribution of the response, which is known as the central subspace (CS). When the response is multivariate, the problem becomes con...
268. Scale-Consistent Posterior Dynamics for Diffusion Inverse Problems ​
Author: Zhaoqiang Liu, Tongyao Pang, Ruibing Wang, Yang Zheng
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2608.15144v1 Announce Type: cross Abstract: Posterior sampling with a pretrained diffusion prior is governed by a conditional score whose intermediate likelihood component is generally intractable. We begin from an ideal one-parameter posterior SDE family in which a stochasticity parameter con...
269. An Adaptive Gradient Clipping and Noise Injection Mechanism for Differentially Private Federated Learning ​
Author: Wenjing Wei, Alla Jammine, Farid Nait-Abdesselam
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.15153v1 Announce Type: cross Abstract: Differentially private federated learning must balance privacy protection against model accuracy and training efficiency. Static gradient clipping applies a fixed threshold throughout training and across model layers, which can cause excessive clippi...
270. Beyond Effective Sample Size: Effective Number of Proposals for Adaptive Importance Sampling ​
Author: Ali Mousavi, Victor Elvira
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.15154v1 Announce Type: cross Abstract: Population-based adaptive importance sampling (AIS) methods use a set of proposal densities to approximate complex target distributions. Their performance is commonly assessed through effective sample size (ESS) and related weight-based diagnostics, ...
271. Identifying parameter couplings and uncertainties of mixed-noise stochastic systems via full-covariance Gaussian mixture network ​
Author: Xiaolong Wang, Xiangwen Hao, Jing Feng, Yuanyuan Liu, Yong Xu
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, physics.comp-ph
arXiv:2608.15198v1 Announce Type: cross Abstract: Parameter identification of stochastic dynamical systems driven by mixed noises is challenging due to intractable likelihood functions. We propose PENN-GMD, a parameter estimation neural network that maps partially observed trajectories to a Gaussian...
272. The Distributional View of Knowledge Distillation ​
Author: Gordei Verbii, Juho Lee
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.15215v1 Announce Type: cross Abstract: Token-level knowledge distillation (KD) matches two conditional distributions per position, yet the standard objectives compare them pointwise: a Kullback-Leibler gradient is blind to which wrong token receives probability mass. We develop a distribu...
273. Decentralized Federated Learning for Heterogeneous Multi-Task Semantic Communication ​
Author: Lin Yin, Tiejun Lv, Weicai Li, Xi Yu, Xiaoyu He
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.15256v1 Announce Type: cross Abstract: Collaborative training in distributed semantic communication (DSC) networks typically relies on decentralized federated learning (DFL). However, pushing topology-agnostic aggregation into heterogeneous, multi-task environments creates a fundamental b...
274. BrainLinear: A Linear Model for Brain Network Analysis in Sparse Tangent Subspaces ​
Author: Sijing Wu, Dongyuan Li, Miaoting Huang, Weiwei Ye, Ying Zhang, Feng Xia, Renhe Jiang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.GR, cs.LG
arXiv:2608.15266v1 Announce Type: cross Abstract: Functional connectome analysis examines brain-region interactions to understand and identify disorders such as autism spectrum disorder and Alzheimer's disease. Existing methods typically use GNNs and Transformers to model the full functional connect...
275. Memory-Bounded Continuation of Greedy Sampling for Continual Anomaly Detection ​
Author: Yoon Gyo Jung, Jaewoo Park, Kuan-Chuan Peng, Seongdeok Bang, Octavia Camps
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.15277v1 Announce Type: cross Abstract: Greedy sampling produces a compact yet representative summary of normal data, which is essential for reliable anomaly detection that relies on measuring distance from normality. For continual anomaly detection where tasks arrive sequentially, extendi...
276. Convolution Smoothed Quantile Regression for XGBoost ​
Author: Mandy Yao (University of Toronto), Meredith Franklin (University of Toronto)
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.15290v1 Announce Type: cross Abstract: The increasing availability of large and complex datasets across many scientific disciplines has led to widespread adoption of machine learning (ML) for prediction. However, most ML algorithms focus on point estimation and provide limited information...
277. A Unified Geometric Framework for Developmental Analysis of Spatial Transcriptomic Data ​
Author: Mary Chriselda Antony Oliver, Kaitlyn Hohmeier, Tuyen Tran, Alejandra Castillo, Caroline Moosm"uller, Shiying Li
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.MG
arXiv:2608.15306v1 Announce Type: cross Abstract: High-throughput single-cell and spatial transcriptomic technologies provide high-resolution snapshots of heterogeneous cellular states, but their destructive nature prevents repeated measurements of the same cells over time. Consequently, temporal an...
278. FedADB: Class Anchor-Driven Dual-Branch Federated Learning for Mitigating Forgetting ​
Author: Zhenyan Liu, Hua Zhang, Haoran Gao, Qi Li, Hongliang Zhu, Huiyu Zhou, Zongliang Shen, Yanxin Xu, Jiahui Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.15310v1 Announce Type: cross Abstract: Multimodal data collected by heterogeneous devices are used for collaborative training, where federated learning (FL) serves as a key paradigm for effective distributed modeling with data privacy preservation. However, local training suffers from the...
279. The Physical Cutoff Does Not Restore Homogenization: Phase-Dependent Burning in the Strain G-Equation ​
Author: Michele Caprio
Published: 8/19/2026, 4:00:00 AM
Categories: math.AP, cs.LG
arXiv:2608.15337v1 Announce Type: cross Abstract: We disprove the expectation stated by Xin, Yu, and Ronney that the physical positive part strain $G$-equation should possess an effective burning velocity in cellular flows. For the standard cellular flow in dimension two $V_A(x_1,x_2)=A(-\sin x_1\co...
280. Prediction Inference of Time Series with Standard ReLU Deep Neural Networks ​
Author: Kejin Wu
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO
arXiv:2608.15362v1 Announce Type: cross Abstract: We propose a methodology based on the standard ReLU Deep Neural Networks (DNN) to make predictions and quantify their uncertainty. Classically, people rely on linear, non-linear, or non-parametric kernel methods to fit and then predict the time serie...
281. FedPA-LoRA: Product-Aligned Framework for Mitigating Aggregation and Initialization Errors in Heterogeneous Federated LoRA ​
Author: Juseok Jeon, Ramy E. Ali, Doyun Kwon, Myungbeom Her, Jinhwi Kim, Jinhyun So
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.15381v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) enables efficient federated fine-tuning of large language models, but its factorized parameterization creates a tension between accurate aggregation of local updates and continuity of locally optimized factors. Factor-wise ...
282. ETHOS: Towards a Modular Ethics Framework for Clinical Multi-Agent Systems ​
Author: Rakesh Sharma, Sydney Pugh, Cameron Beeche, Pankhuri Singhal, Rachel Wu, Margaret Eby, Jeffrey Duda, James Gee, Kyra O'Brien, Hersh Sagreiya, Marina Serper, Victoria Gershuni, Angela Bradbury, Anurag Verma, Eric Eaton, Kevin B. Johnson, Walter Witschey
Published: 8/19/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.LG
arXiv:2608.15424v1 Announce Type: cross Abstract: The rapid adoption of large language models has enabled the development of clinical multi-agent systems (MAS) capable of integrating multimodal patient data and supporting increasingly complex clinical decision-making. However, the deployment of thes...
283. NeuRoute: Logit-Guided Neural Routing for Billion-Scale Vector Search with Sub-Hour Index Construction ​
Author: Xingqiao Wang, Zi Wang, Xiaowei Xu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DB, cs.IR, cs.LG
arXiv:2608.15438v1 Announce Type: cross Abstract: Building approximate nearest neighbor (ANN) indexes at billion scale is often dominated by expensive global clustering or graph construction, making time-to-index a first-order systems concern. We present NeuRoute, a learned hashing index that turns ...
284. Language models suffer from a curse of ambiguity ​
Author: Nicolas Zucchet, Hyun Dong Lee, Scott Linderman
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.NE
arXiv:2608.15448v1 Announce Type: cross Abstract: Large language models increasingly rely on sampling as a driver of their own improvement, making the fidelity of their learned distributions more critical than ever. Yet, not all distributions are equally easy to learn. In this work, we identify a cu...
285. Maintaining IoT Device Identification under Concept Drift via Budget-Aware Traffic Labeling ​
Author: Shayan Azizi, Norihiro Okui, Masataka Nakahara, Ayumu Kubota, Gustavo Batista, Hassan Habibi Gharakaheili
Published: 8/19/2026, 4:00:00 AM
Categories: cs.NI, cs.CR, cs.LG
arXiv:2608.15465v1 Announce Type: cross Abstract: Identification of IoT device types from passive traffic is increasingly used for security management in enterprise and ISP networks. However, the performance of machine learning-based classifiers gradually degrades under concept drift as device behav...
286. Do Language Models Consistently Encode the Current Year? ​
Author: Suze van Adrichem, Aditi Bhaskar, Diyi Yang, Christopher Potts, Jing Huang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.15507v1 Announce Type: cross Abstract: A consistent concept of the current time is important for temporal reasoning, yet how language models represent the current time is not well understood. We contribute two tasks that probe the current year in conceptually distinct ways: an associative...
287. Temporal Logic Guided Universal Task Representations for Reinforcement Learning ​
Author: Hao Zhang, Zhangli Zhou, Zhen Kan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.RO, cs.FL, cs.LG
arXiv:2608.15509v1 Announce Type: cross Abstract: Task guided agents demonstrate strong performance in a wide range of complex tasks. However, most existing task representation algorithms are tailored to specific contexts and struggle to generalize across diverse scenarios. Moreover, they typically ...
288. DeltaLog: Deferred Materialization of Recurrent States for Linear Attention Decoding ​
Author: Junqing Lin, Jingwei Sun, Guangzhong Sun
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2608.15533v1 Announce Type: cross Abstract: Linear attention models eliminate the quadratic prefix computation and context-growing KV cache of softmax attention by replacing pairwise token interactions with recurrent state updates. However, existing decoding implementations often materialize a...
289. L3Cube-IndicQuest v2: A Large-Scale Multilingual Benchmark for Evaluating Factual Knowledge of Large Language Models Across Indic Languages ​
Author: Rinit Jain, Tirthraj Mahajan, Advait Joshi, Raviraj Joshi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.15535v1 Announce Type: cross Abstract: We present L3Cube-IndicQuest v2, a large-scale gold-standard multilingual question-answering benchmark for evaluating the India-specific factual knowledge of Large Language Models (LLMs). The benchmark comprises 3,471 curriculum-grounded English ques...
290. RigidBench: Evaluating Rigid-Body Physics in Video Generation Models ​
Author: Swarnim Jain, Shangzhe Wu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.15555v1 Announce Type: cross Abstract: Video models are increasingly used to predict what happens next in a scene, yet the metrics commonly used to compare their outputs say little about whether the predicted objects move correctly. Motion, geometry, identity, background stability, and vi...
291. A Counterexample to the Tang Zhang Schatten Norm Conjecture and Sharp Positive Results ​
Author: Zijian Zeng, Houde Liu, Kurunathan Ratnavelu
Published: 8/19/2026, 4:00:00 AM
Categories: math.CO, cs.LG, math.FA
arXiv:2608.15558v1 Announce Type: cross Abstract: For $m\geq 2$, let $c_p(m)$ be the all-dimensional best constant in $$ \left|\sum_{k=1}^m A_k\right|p \leq c_p(m)\left|\sum^m |A_k|\right|_p. $$ Tang and Zhang conjectured an explicit formula for every finite $p>1$. We disprove the conject...
292. Rotation-Invariant Multi-IMU Activity Recognition under Independent Per-Location Orientation Shifts ​
Author: Seungyeol Baek, Yoonbyung Chai, Yonghyeon Lee, Sungjoon Choi, Sungho Suh
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.15621v1 Announce Type: cross Abstract: Human Activity Recognition (HAR) with self-administered wearables, such as at-home rehabilitation and exercise monitoring, often requires reattaching inertial measurement units (IMUs) across sessions. In multi-IMU settings, this can induce independen...
293. When Is Shallow Enough? Adaptive Split Federated Learning with Client-Specific Sufficiency Estimation ​
Author: Wenhao Yuan, Chenchen Lin, Wenhao Hu, Jian Chen, Jinfeng Xu, Shujie Li, Edith Cheuk Han Ngai
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG
arXiv:2608.15639v1 Announce Type: cross Abstract: \textit{Split Federated Learning} (SFL) enables distributed model training by splitting networks between the server and clients. However, under client heterogeneity, the conventional static split strategy may be suboptimal because clients can differ ...
294. On Stopping Rules and Spatial Adaptation for CART ​
Author: Zineng Xu, Yuchao Cai, Yan Shuo Tan
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2608.15649v1 Announce Type: cross Abstract: The popular CART algorithm for regression trees combines a greedy splitting rule with a stopping rule, but while the splitting rule has been well studied, the statistical role of stopping rules is less well understood. Meanwhile, although regression ...
295. Adaptive Heterogeneous Compression for Resource-Efficient Federated Knowledge Distillation ​
Author: Chenwang Liu, Yijun Liu, Chang Liu, Xu Zhang, Pengchao Han
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2608.15660v1 Announce Type: cross Abstract: Federated learning (FL) enables privacy-preserving distributed model training but faces challenges from heterogeneous model architectures and limited communication resources at the network edge. Federated knowledge distillation (FedKD) alleviates mod...
296. Integrating Persuasion Theory into the Epidemiological Modelling of Health Misinformation Spread on Social Media ​
Author: Mkululi Sikosana, Sean Maudsley-Barton, Oluwaseun Ajao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SI, cs.AI, cs.CL, cs.LG
arXiv:2608.15689v1 Announce Type: cross Abstract: This study presents a hybrid epidemiological and behavioural framework to simulate the spread of health misinformation on social media. We extend the classical Susceptible--Infected--Recovered (SIR) model to a six-compartment structure (SIRMMM), inco...
297. Adding Voice Cloning to Text-to-Audio-Video Models with a Single Zero-Initialised Layer ​
Author: Ivan Mikheev, Viacheslav Vasilev, Anna Dmitrienko, Alexey Letunovskiy, Ivan Kirillov, Kirill Chernyshev, Denis Dimitrov
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG, cs.MM
arXiv:2608.15690v1 Announce Type: cross Abstract: Text-to-audio-video (T2AV) generation models produce a video and its soundtrack from a textual description, but offer no control over whose voice speaks in the output. We show that a base T2AV model can be turned into a voice-cloning model by adding ...
298. BERTopic-Virality Prioritisation: A Scalable Framework for Thematic and Comparative Analysis of COVID-19 and Monkeypox Misinformation on Twitter ​
Author: Mkululi Sikosana, Sean Maudsley-Barton, Oluwaseun Ajao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SI
arXiv:2608.15691v1 Announce Type: cross Abstract: Health misinformation circulating during pandemics can gain traction rapidly, creating harmful narratives that compete with public health guidance. Most topic-modelling pipelines treat engagement as an external outcome, limiting their ability to prio...
299. Large Models for Small Devices: Recent Advances and Empirical Analysis of Edge AI Deployment ​
Author: Subhransu Das, Jiaming Cheng, Arnav Kumar, Sadia Afrose, Mingzhe Han, Michael Silagy, Shreya Palande, Brijesh Soni, Rajiv Ramnath
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.15693v1 Announce Type: cross Abstract: Running large AI models on resource-constrained edge devices requires model compression to reduce model size and computation. What compresses well, however, need not deploy well. We survey dozens of recent works that report compression results on rea...
300. Deep learning-based computed tomography (CT) derived body composition classifier for colorectal cancer patients ​
Author: Eve Harling (James Watt School of Engineering, College of Science & Engineering, University of Glasgow, Glasgow, UK), Chattarin Pumtako (Academic Unit of Surgery, School of Medicine, College of Medical Veterinary & Life Sciences, University of Glasgow, Glasgow, UK), Bernd Porr (James Watt School of Engineering, College of Science & Engineering, University of Glasgow, Glasgow, UK), Donald C McMillan (Academic Unit of Surgery, School of Medicine, College of Medical Veterinary & Life Sciences, University of Glasgow, Glasgow, UK), Ross D Dolan (Academic Unit of Surgery, School of Medicine, College of Medical Veterinary & Life Sciences, University of Glasgow, Glasgow, UK)
Published: 8/19/2026, 4:00:00 AM
Categories: eess.IV, cs.LG
arXiv:2608.15712v1 Announce Type: cross Abstract: Background: Accurate body composition analysis using Computed Tomography (CT) scans is essential for assessing skeletal muscle area (SMA) and skeletal muscle density (SMD), key markers of nutritional status in cancer patients. Conventional manual met...
301. Continuous Quantum Feedback Control via Kraus-Parameterized Belief Reinforcement Learning ​
Author: Priyanshi Singh, Krishna Bhatia
Published: 8/19/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.15715v1 Announce Type: cross Abstract: Quantum feedback control requires acting on noisy continuous measurement records without direct access to the underlying quantum state. We propose Kraus-Parameterized Belief Reinforcement Learning, a pipeline in which a recurrent encoder, constrained...
302. PLeDO: Pain Level Detection for Osteoarthritis from EMR Data ​
Author: Yuhao Chen, Jiahao Cai, Nafiz Sadman, Farhana Zulkernine, John Queenan, David Barber
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.ET, cs.IR, cs.LG
arXiv:2608.15719v1 Announce Type: cross Abstract: Osteoarthritis (OA) is a progressive chronic joint disease resulting in a breakdown of articular cartilage and bone when damaged joint tissues are not able to normally repair themselves. The aim of this pilot research study is to understand the pain ...
303. Machine Learning Approaches to Decoding Topological Quantum Codes ​
Author: Changwon Lee, Tak Hur, Jeongwoo Jae, Daniel K. Park
Published: 8/19/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.15760v1 Announce Type: cross Abstract: Decoding is an essential component of quantum error correction (QEC), translating stabilizer measurement outcomes into corrective actions that suppress logical errors and preserve logical quantum information. Building fault-tolerant architectures req...
304. Provenance, Not Behaviour: A Serialisation Artifact in Edge-IIoTset and a Leakage-Free Benchmark for Precision-Agriculture Intrusion Detection ​
Author: Mostafa M. Galal
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.15761v1 Announce Type: cross Abstract: Edge-IIoTset is the reference benchmark for machine-learning intrusion detection in the industrial Internet of Things, and results reported on it cluster above 99%. We show that much of that performance is not intrusion detection. The preprocessing r...
305. Global Simulation-Guided Dynamic Operator Scheduling for Efficient Multi-Tenant Model Serving ​
Author: Weinan Liu, Zeyuan Ding, Dian Ding, Chengcheng Wan, Lu Tang, Guangtao Xue, Jiwu Shu, Yiming Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.OS, cs.DC, cs.LG
arXiv:2608.15762v1 Announce Type: cross Abstract: Container-granularity scheduling leaves abundant short-lived idle slices within containers unexploited. Reallocating containers is too heavyweight to utilize such fine-grained opportunities under SLA constraints, and operator-level scheduling require...
306. Inferential Evaluation of Surrogate-Derived Models under Covariate Shift ​
Author: Longtian Shi, Molei Liu, Doudou Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP, stat.ME
arXiv:2608.15783v1 Announce Type: cross Abstract: In transfer-learning settings, a model derived from abundant surrogate labels may be deployed in a target population where gold-standard outcomes are unobserved. Evaluating its target performance is essential for determining whether decisions based o...
307. PWLR: Pairwise Witness Local Rejection for Boundary-Aware Out-of-Distribution Detection ​
Author: Chengyao Jia, Ruixuan Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.15802v1 Announce Type: cross Abstract: Out-of-distribution (OOD) detection remains challenging for image classifiers, especially when near-OOD samples lie close to in-distribution (ID) class boundaries. Recent vision-language detectors improve OOD detection through class semantics, local ...
308. QuantumPhaseNet: A Gauge-Covariant Geometric and Quantum-Spectral Theory of Semantic Concept Hierarchies with Prototype Validation of a Classical Quantum-Inspired Model ​
Author: Kiyotaka Kasubuchi, Kazuo Fukiya
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.15820v1 Announce Type: cross Abstract: We present QuantumPhaseNet, a gauge-covariant geometric and quantum-spectral extension of Transformer representations. Context-dependent semantic states are modeled as complex amplitudes; a covariant phase rate induces a semantic wavelength used as a...
309. How Many Samples Are Needed to Determine Causal Direction? Sharp Minimax Bounds for Bivariate LiNGAM ​
Author: Jikai Jin
Published: 8/19/2026, 4:00:00 AM
Categories: math.ST, cs.LG, econ.EM, stat.ML, stat.TH
arXiv:2608.15840v1 Announce Type: cross Abstract: We study how many observations are needed to determine the causal direction between two linearly related variables. Classical LiNGAM theory shows that independent non-Gaussian disturbances identify the direction, but does not quantify the difficulty ...
310. Generalized Linear Bandits with Memory ​
Author: Heesang Ann, Hyunjun Choi, Taehyun Hwang, Younghoon Shin, Haeju Cheong, Min-hwan Oh
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.15848v1 Announce Type: cross Abstract: We study generalized linear bandits with memory, an endogenous non-stationary setting in which rewards depend on past actions through a finite memory matrix. Building on prior work for linear models (Clerici et al., 2024), we show that the previously...
311. CoupVisor: Strategy Optimization by Round and Challenge Decision Support ​
Author: Cris Huynh
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.GT, cs.LG
arXiv:2608.15868v1 Announce Type: cross Abstract: This paper presents CoupVisor, a decision-support system for the hidden-information card game Coup. It addresses two questions: what a player should do on each turn, and when a player should challenge an opponent's claim. The system is built around a...
312. Beyond Visual CoT: Internalized Visual Thinking for Proactive Video Reasoning ​
Author: Xiaoyu Zhu, Xinke Deng, Suresh Taddewadikar, Arnab Kumar Mondal, Zhongyu Jiang, Ian Fasel, Joerg Liebelt
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG, cs.MM
arXiv:2608.15869v1 Announce Type: cross Abstract: Multimodal large language models increasingly use visual chain-of-thought (Visual CoT) to reason about spatial, temporal, and embodied environments. By generating intermediate reasoning images, Visual CoT provides an intuitive mechanism for visual fo...
313. Large Language Models as Implicit Sociological Models: Reconstructing Voting Behaviour from Sociodemographic Profiles ​
Author: Roman Neruda, Martin Bako\v{s}, Josef \v{S}lerka, V'it Tu\v{c}ek, Petra Vidnerov'a, Gabriela Kadlecov'a
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CY, cs.CL, cs.LG
arXiv:2608.15871v1 Announce Type: cross Abstract: Large language models (LLMs) trained on large-scale internet corpora encode extensive statistical regularities about social identities, attitudes, and political behaviour. This paper introduces and evaluates a methodological framework that leverages ...
314. Resource-Efficient QUBO Formulation for Anchored Currency Arbitrage ​
Author: Eric A. F. Reinhardt, Adam J. Hauser
Published: 8/19/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.15889v1 Announce Type: cross Abstract: Currency arbitrage (CA) involves trading currencies in cycles to exploit discrepancies in market valuations. Quadratic unconstrained binary optimization (QUBO) involves minimizing a quadratic cost (energy) function of binary variables. Previous works...
315. Crystal-structure design by agentic AI in a language of motifs ​
Author: Dinh-Khiet Le, Minh-Quyet Ha, Hong-Phuc Vu-Dinh, Takashi Miyake, Hiori Kino, Hieu-Chi Dam
Published: 8/19/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG
arXiv:2608.15900v1 Announce Type: cross Abstract: Data-driven materials discovery interpolates more reliably than it extrapolates and seldom reaches new structure types. We present MatEvolve, an agentic-AI framework designing crystals, proposing each candidate with a stated rationale and testing it....
316. $S^3$: A Smooth Simulation Surrogate for Optimizing Discrete Abstractions of Dynamical Systems ​
Author: Jordan Peper, James Mathias Gast, Vignesh Nanduri, Tanmayee Maram, Ethan Howes, Ivan Ruchkin
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2608.15920v1 Announce Type: cross Abstract: Intelligent systems are increasingly deployed in safety-critical settings with black-box controllers, including neural networks. The properties and behaviors of these end-to-end systems can be studied with abstraction-based methods that replace them ...
317. The Null Token Knows: Reducing Message-Free Hallucination in ASR and NMT ​
Author: Kirill Borodin, Vasiliy Kudryavtsev, Ivan Viakhirev, Grach Mkrtchian
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SD
arXiv:2608.15940v2 Announce Type: cross Abstract: Modern encoder-decoder systems can produce fluent text even when their input contains no recoverable message. We study this failure in ASR and NMT through the models' reserved null tokens, asking whether the score for ending generation already carrie...
318. Ask to Be Sure: Informative Interactions for Confident Multi-Turn LLM Recommendation ​
Author: Cedar Site Bai, Duanshun Li, Zhenyu Liao, Sheikh Sarwar, Huiyuan Chen, Yuan Chen, Changhe Yuan, Haiyang Zhang, Qilin Qi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG
arXiv:2608.15949v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled their use as conversational recommender systems (CRS), demonstrating strong recommendation accuracy and natural dialogue. However, guiding multi-turn interactions to elicit user preferences...
319. Functional anatomy of Pythia-Herwig differences with Kolmogorov-Arnold networks ​
Author: Arghya Chattopadhyay
Published: 8/19/2026, 4:00:00 AM
Categories: hep-ph, cs.LG
arXiv:2608.15952v1 Announce Type: cross Abstract: Differences between high-energy event generators can arise at several stages of the collision simulation, from the hard scattering through parton showering and hadronization to the final event. These differences are usually summarized using observabl...
320. Solvable Sokoban Without a Solver via Diffusion ​
Author: Sina Baghal
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.GT, cs.LG
arXiv:2608.15958v1 Announce Type: cross Abstract: Deciding whether a Sokoban puzzle is solvable is PSPACE-complete (Culberson, 1997): solutions can be exponentially long and there is no short certificate to check. Solvability is also a fragile property, since even a single misplaced wall can silentl...
321. A Scalable Pipeline for LLM-Teacher Distillation Labeling: Work-Stealing Job Scheduling and Memory-Aware GPU Concurrency ​
Author: Ravi Satya Durga Prasad Yenugula
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.CL, cs.LG
arXiv:2608.15975v1 Announce Type: cross Abstract: Labeling large text corpora with LLM teachers has become a practical route to training data at scale. At millions of items, hand-labeling every batch is not feasible, and two questions dominate: what label quality a teacher buys per dollar, and how t...
322. Whose Gold? Annotator-Pool Disagreement Is Large at the Item Level, and Hidden by Small Leaderboards ​
Author: Anik Jha
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.15980v1 Announce Type: cross Abstract: Preference benchmarks are built by hiring annotators, and the identity of those annotators is treated as an implementation detail. We measure what that detail buys. On the 2,885 MultiPref items where both pools are internally unanimous, so no tie-bre...
323. Learning Varying Physical Therapist-Patient Interactions for Robot-mediated Upper Limb Task-Specific Training ​
Author: Jia Quan Loh (Human Robotics Laboratory, Department of Mechanical Engineering, The University of Melbourne), Vincent Crocher (Human Robotics Laboratory, Department of Mechanical Engineering, The University of Melbourne), Marlena Klaic (Melbourne School of Health Sciences, The University of Melbourne), Denny Oetomo (Human Robotics Laboratory, Department of Mechanical Engineering, The University of Melbourne), Ying Tan (Human Robotics Laboratory, Department of Mechanical Engineering, The University of Melbourne)
Published: 8/19/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.15995v1 Announce Type: cross Abstract: Upper extremity motor function recovery is positively linked to Task-Specific Training (TST) and sufficient therapy dosage. Rehabilitation robots can increase TST dosage via controlled, repetitive treatment and free therapists to simultaneously manag...
324. Structured Prediction for Scalable Spreadsheet Table Understanding: From Cell Types to Table Ranges (Extended Version) ​
Author: Antoine Gauquier, Ioana Manolescu, Pierre Senellart
Published: 8/19/2026, 4:00:00 AM
Categories: cs.IR, cs.DB, cs.LG
arXiv:2608.16050v1 Announce Type: cross Abstract: Spreadsheets are a primary medium for publishing tabular data, yet automatically extracting structured content from them remains difficult due to heterogeneous layouts, diverse file formats, and inconsistent organizational conventions. We address two...
325. GOD: Enhancing Generalization via Deep Grafting for Sequential Recommendation ​
Author: WooJoo Kim, JunYoung Kim, JaeHyung Lim, HwanJo Yu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.16073v1 Announce Type: cross Abstract: Sequential recommenders often struggle with sparse and noisy histories, limiting generalization to unseen interactions. Knowledge distillation mitigates this by transferring dense supervision from a teacher to a student. However, most distillation me...
326. Eigenanalysis framework for autoregressive neural emulators of multi-scale chaotic dynamics ​
Author: Conrad Ainslie, Pedram Hassanzadeh, Michael W. Mahoney, Ashesh Chattopadhyay
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, nlin.CD, physics.comp-ph
arXiv:2608.16084v1 Announce Type: cross Abstract: Neural autoregressive models have rapidly emerged as powerful emulators of high-dimensional chaotic systems, yet their long-term instability and error growth remain poorly understood, leading to ad-hoc solutions. Here, we develop an eigenanalysis fra...
327. Representation Is Not Enough: Body-Localized Thermal Evidence for Contactless Stress and Craving Sensing in Opioid Use Disorder ​
Author: Sachin Deb, Harshit Sharma, Asif Salekin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.16087v1 Announce Type: cross Abstract: Removing wearables from physiological monitoring also removes their supervision: the signal indicating where and when a stress response occurred. Contactless stress sensing therefore becomes a weakly supervised evidence-localization problem, where a ...
328. Protein Structure Prediction: From Evolutionary Constraints to Generative Modeling ​
Author: Wengan He, Yongsheng Luo, Lihong Jiang, Wenhui Xu, Yu Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.16094v1 Announce Type: cross Abstract: Accurate protein structure prediction is fundamental to structural biology because protein structure underlies molecular function and provides a basis for mechanistic interpretation. Recent advances in deep learning have transformed the field from mu...
329. EMS Coreset: An Efficient Expectation-Maximization Algorithm for Sinkhorn Coreset ​
Author: Haoyun Yin, Chuanhui Liu, Xiao Wang
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.16101v1 Announce Type: cross Abstract: Coresets distill large datasets into small, representative subsets for efficient downstream learning. Yet Optimal Transport (OT)-based selection typically requires intensive computation of transport plans, limiting scalability. We introduce a scalabl...
330. TokenSTFormer: A Tokenized Spatial-temporal Attention Model for Holistic Motion Analysis in Adolescent Idiopathic Scoliosis Screening ​
Author: Dong Chen, Kenneth M. C. Cheung
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.16122v1 Announce Type: cross Abstract: Adolescent Idiopathic Scoliosis (AIS) is a prevalent spinal deformity in adolescents that, if left untreated, can result in severe health outcomes. Traditional screening methods are limited by subjective interpretation, reliance on professional exper...
331. Coded Hankel Polynomial Chaos: Spectral Identification of Dominant Polynomial-Chaos Modes ​
Author: Zhiliang Deng, Xiaomei Yang
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.16126v1 Announce Type: cross Abstract: Identification of dominant polynomial-chaos modes is usually formulated as a sparse-regression problem on a sampled multivariate polynomial dictionary. We develop coded Hankel polynomial chaos (CH-PC), a complementary spectral formulation for dominan...
332. When Single-Dataset Conclusions Fail: A 45-Task Study of Threshold Tuning and Resampling for Imbalanced Classification ​
Author: Diyorbek Musaev
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.16147v1 Announce Type: cross Abstract: Class-imbalance handling is routinely evaluated on a single benchmark dataset, and the resulting conclusions are reported as if they were properties of the method. We show this practice is unsafe. On the public Kaggle credit-card fraud dataset, under...
333. Asymptotics-guided learning and symbolic regression for dispersive resonances ​
Author: Konstantinos Alexopoulos, Josselin Garnier
Published: 8/19/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math-ph, math.AP, math.MP, physics.optics
arXiv:2608.16152v1 Announce Type: cross Abstract: We study resonance prediction in dispersive media, formulated as nonlinear spectral problems for volume integral operators. The main idea is to use asymptotic analysis not only as a baseline approximation, but also as a guide for constructing predict...
334. Digital Twin Degradation: Detecting Cyber Physical Attacks via Temporal Inconsistencies ​
Author: Konstantinos E. Kampourakis, Vasileios Gkioulos, Sokratis Katsikas
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2608.16159v1 Announce Type: cross Abstract: Digital Twins (DTs) are increasingly used to monitor and analyze Cyber Physical Systems (CPS). However, in adversarial environments, the fidelity of a DT cannot be assumed. Communication delays, data manipulation, sensor degradation, or partial infor...
335. Domain-Specific Text Embedding Models for Entity Resolution ​
Author: Khajesh Sapram, Srivardhani Raju, Kishore Konda
Published: 8/19/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG
arXiv:2608.16161v1 Announce Type: cross Abstract: General-purpose text embedding models are designed to capture semantic similarity but are not optimised for distinguishing entity records that represent the same real-world business or person. This limitation affects applications such as entity resol...
336. RadioVIL: Anomaly-Aware Diffusion Models for Radio Map Inpainting and Zero-Shot Vehicle Localization ​
Author: Ruixin Zhao, Xiucheng Wang, Qiming Zhang, Nan Cheng, Ruijin Sun, Conghao Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.16167v1 Announce Type: cross Abstract: High-precision radio map construction is essential for emerging 6G Integrated Sensing and Communication (ISAC) applications, including digital twins and intelligent transportation. However, existing deep learning methods predominantly treat this as a...
337. Decorrelation Is Not Complementarity: Skill, Not Lineage, Governs Trusted-Monitor Ensembles ​
Author: Anik Jha
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.16190v1 Announce Type: cross Abstract: Trusted monitoring has a cheap, trusted model score a stronger untrusted model's actions, and a diverse ensemble of them beats a single stronger monitor at matched cost. They are built by minimising average pairwise correlation, and that paper's twel...
338. Convolution-Free Holistic Multivariance Decomposition Layer for Efficient Hyperspectral Image Classification Tensor Networks ​
Author: S"uha Tuna, "Ulker Ba\c{s}ar
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, math.OC
arXiv:2608.16241v1 Announce Type: cross Abstract: Feature extraction for hyperspectral image classification is conventionally addressed using rigid tensor decompositions that fail to capture complex spatio-spectral interdependencies, or heavily parameterized convolutional neural networks that are co...
339. CoM$^3$eT: A foundation model for medical image analysis through federated, multidimensional context integration ​
Author: J. Raphael Sch"afer, Kai Geissler, Till Nicke, Chiara Tappermann, Karoline Heber, Eike Petersen, Habib Mergan, Lars Ole Schwen, Nick Weiss, Annika Gerken, Jan Hendrik Moltz, Tom Bisson, Isil Dogan O, Tim-Rasmus Kiehl, Norman Zerbe, Sefer Elezkurtaj, Robin S. Mayer, Nadine Flinner, Peter Wild, Isabel Dahm, Felix Peisen, Heinrich von Busch, Robert Grimm, Sebastian Arndt, Lisa Siegler, Matthias Stefan May, Antje Prasse, Natalia Artysh, Fabian Kiessling, Johannes Lotz
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.16268v1 Announce Type: cross Abstract: Medical foundation models improve generalization when training AI models with limited labeled data, but remain confined to a single specialty, such as pathology or radiology, and to either sparse or dense outputs, such as classification or segmentati...
340. Domain-Agnostic Neural Topic Modeling with Contextual Token-Level Semantic Graph Representation ​
Author: Seung-Won Seo, Won Ik Cho, Yongmin Yoo
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.16269v1 Announce Type: cross Abstract: Recent advances in neural topic models with pre-trained language models (PLMs) have achieved strong performance by leveraging general-domain pre-training, yet their topic interpretability often degrades on specialized corpora. This limitation primari...
341. Correlation Clustering with Random Partial Information ​
Author: Rajath Rao K. N., Jens Schl"oter, Sami Davies, Amira Ouchene, Yasamin Nazari
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DS, cs.LG
arXiv:2608.16315v1 Announce Type: cross Abstract: Correlation clustering is a fundamental unsupervised learning problem. On complete graphs, both the min-disagreement and min-max objectives admit constant-factor approximations, yet on general (non-complete) graphs, the best guarantees blow up to $O(...
342. Predicting, Evaluating, and Explaining Top Misinformation Spreaders via Archetypal User Behavior ​
Author: Enrico Verdolotti, Luca Luceri, Silvia Giordano
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SI, cs.CY, cs.LG
arXiv:2608.16323v1 Announce Type: cross Abstract: The spread of misinformation on social networks poses a significant challenge to online communities and society at large. Not all users contribute equally to this phenomenon: a small number of highly effective individuals can exert outsized influence...
343. LaGSplat: Inferring Physics-Governed Interactive Simulation from Monocular Video Using Latent Lagrangian Gaussian Splatting ​
Author: Louen Pottier
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.16324v1 Announce Type: cross Abstract: We present LaGSplat (Latent Lagrangian Gaussian Splatting), a framework that infers interactive, physics-governed dynamics from one or a few monocular videos. At inference it lets a user push on the filmed object, rigid or deformable, with an externa...
344. Beyond Binary Priorities: Multi-Tier SLA Scheduling for Large Language Model Serving ​
Author: Anders Vestrum, Arya Raeesi, Hanna Roed
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AR, cs.DC, cs.LG
arXiv:2608.16336v1 Announce Type: cross Abstract: Modern LLM serving deployments must simultaneously satisfy heterogeneous service-level objectives (SLOs) across a diverse population of user tiers, ranging from latency-critical API calls to background batch processing. Llumnix introduced a dynamic, ...
345. LiD-GLM: Lipschitz-constrained Deep Generalized Linear Models ​
Author: Tom Splittgerber, Niklas Koenen, Marvin N. Wright, Werner Brannath
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.16340v1 Announce Type: cross Abstract: The combination of traditional statistical models and neural network (NN) components into semi-structured hybrid models is an intriguing approach to construct models that, ideally, combine traditional interpretability with the unprecedented flexibili...
346. Architecture-Dependent Causal Transfer of Activation States Across Large Language Models ​
Author: Fernando Cardenas Piepereit
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.16347v1 Announce Type: cross Abstract: Direct communication between AI systems relies on natural language as an intermediate layer, incurring encoding/decoding overhead, token cost, and latency. We ask whether internal activation states can instead be transferred causally between differen...
347. Self-Routed Tensor Adapters for Parameter-Efficient Universal Visual Adaptation ​
Author: Suraj Yadav
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.16384v1 Announce Type: cross Abstract: Universal visual representations require adaptation mechanisms that adapt across heterogeneous domains without fragmenting knowledge into domain-specific modules. Parameter-efficient fine-tuning adapts frozen visual foundation models efficiently, but...
348. Mint-Agent: Introducing Finance-Native Agentic Foundation Models ​
Author: Agent Team, B. Zhang, Yaze Geng, Lei Tang, Yaoyang Yi, Zonghan Wu, Yifan Hu, Kun Wang, Qingsong Wen, Yilei Shao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.16386v1 Announce Type: cross Abstract: Financial agents must do more than recall domain knowledge: they must be both reliable, executing precise operations over grounded evidence, and executive, sustaining long-horizon research whose conclusions remain auditable. We present Mint-Agent, a ...
349. POI Recommendation with LLM-Augmented Multi-Graph Learning and Contrastive Alignment ​
Author: Burak Tamer, Wolfram H"opken, Zehui Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.16407v1 Announce Type: cross Abstract: Point-of-interest (POI) recommendation models based on graph neural networks achieve strong performance by propagating collaborative signals over user-item interactions, yet they struggle with the cold-start problem, where items with few or no intera...
350. Self-Supervised Noise2Noise-Enhanced Denoising for Continuous-Scan Air-Plasma THz Spectroscopy ​
Author: Adam Umra, Oways Alsoloh, Oliver Nagy, Aydin Sezgin, Clara Saraceno
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.16454v1 Announce Type: cross Abstract: Terahertz time-domain spectroscopy (THz-TDS) based on air-plasma generation and balanced air-biased coherent detection offers gap-free broadband coverage, but individual continuous-scan traces are strongly affected by pulse-to-pulse fluctuations and ...
351. Offline Reinforcement Learning for Hemodynamic Management of Sepsis in the ICU: a MIMIC-IV Study with Dual Off-Policy Evaluation ​
Author: Marc P'erez-Roig, David Fern'andez-Narro, Carlos S'aez
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.16482v1 Announce Type: cross Abstract: The dosing of intravenous fluids and vasopressors in sepsis is a sequential decision made under uncertainty and guided largely by clinical judgment, which makes it a natural target for reinforcement learning from historical care. Because a learned po...
352. Improved Regret Analysis for Parallel Gaussian Process Bandit Optimization ​
Author: Shion Takeno, Shogo Iwazaki
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.16492v1 Announce Type: cross Abstract: This paper studies the regret analysis for parallel Gaussian process (GP) bandit optimization. The known regret upper bounds for the widely used GP batched upper confidence bound and GP batched Thompson sampling (GP-BTS) suffer from a multiplicative ...
353. Density-Reweighted Entropic Optimal Transport: Decoupling Geometry from Sampling Density ​
Author: Keyi Li, Yuval Kluger, Boris Landa
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP, stat.ME
arXiv:2608.16506v1 Announce Type: cross Abstract: Dataset alignment is a central step in data analysis across science and engineering, where the goal is to match observations between datasets. Entropic Optimal Transport (EOT) offers a computationally tractable framework for this task by encoding cro...
354. LLMs for Zero-Shot Threat Detection via Structured Risk Indicators ​
Author: Abdullah Alghamdi, Siamak Layeghy, Marius Portmann
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.NI
arXiv:2608.16508v1 Announce Type: cross Abstract: We propose a two-stage large language model (LLM) framework for zero-shot detection of insider threats and advanced persistent threats (APTs) from heterogeneous security logs. The framework models user activity as chronological timelines and incorpor...
355. Data-Driven Reconstruction of Spatially Resolved Electron and Ion Energy Distributions from Macroscopic Plasma Quantities with Deep Neural Networks ​
Author: Libin Varghese, Kaushik Prajapati, Bhaskar Chaudhury
Published: 8/19/2026, 4:00:00 AM
Categories: physics.plasm-ph, cs.LG, physics.comp-ph
arXiv:2608.16519v1 Announce Type: cross Abstract: Spatially resolved EEDFs/IEDFs provide essential kinetic information about low-temperature plasmas (LTPs) and play a central role in determining transport, chemical reaction rates, and plasma surface interactions. While kinetic simulations directly r...
356. Automating Learner Assessment: Benchmarking Machine Learning and Deep Learning Models for EEG-Based Familiarity Prediction ​
Author: Isuru Nanayakkara, Thilina Halloluwa
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.HC, cs.LG
arXiv:2608.16541v1 Announce Type: cross Abstract: Objective assessment of learning remains a fundamental challenge in education. Electroencephalography (EEG) provides a direct, non-invasive window into the neural correlates of knowledge acquisition, including cognitive familiarity. This study benchm...
357. Supervising the Path to Fine Scales: GalerkinFlow for Scientific-Field and Image Super-Resolution ​
Author: Zikang Zhan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.CE, cs.LG
arXiv:2608.16546v1 Announce Type: cross Abstract: Most super-resolution models learn from paired data by supervising only the final high-resolution output. This provides little control over how the prediction should evolve between the downsampled observation and its fine target. We introduce Galerki...
358. Probabilistic Circuits as Reasoning Machines in Artificial Intelligence (Part I) ​
Author: Robert Peharz
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, math.PR
arXiv:2608.16565v1 Announce Type: cross Abstract: This cumulative habilitation thesis studies probabilistic circuits (PCs) as a powerful and tractable framework for reasoning and learning under uncertainty in artificial intelligence (AI). It first advocates for probability as a core language for AI,...
359. Degradation-Aligned Self-Supervised Learning for State of Health Estimation of Lithium-Ion Batteries under Label Sparsity ​
Author: Jiaqi Yao, Julia Kowal
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG
arXiv:2608.16612v1 Announce Type: cross Abstract: An accurate estimation of the state of health (SOH) underpins a safe and optimized use of the battery system. Although compelling, data-driven SOH estimation models typically require large amounts of high-quality labeled cycling data, while in practi...
360. The Working Set of a Coding Agent: Coherence Debt in Repository-Scale Tasks ​
Author: Bardia Mohammadi, Lars Klein, Aman Chadha, Akhil Arora, Laurent Bindschaedler
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2608.16630v1 Announce Type: cross Abstract: Repository-scale coding requires an agent to keep tests, imports, configuration, and migration rules consistent within a bounded context window. We model this as reconstructing a coupled-fact graph: at each edit, a required fact comes from recent con...
361. Toward Better Assessment of LLMs' Performance in Clinical Error Detection ​
Author: Yifan Zhang, Rahmatollah Beheshti
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.16643v1 Announce Type: cross Abstract: Automated detection of errors in clinical documentation is a promising application of large language models (LLMs), yet decisions to deploy such models rest on benchmarks that evaluate each clinical note in isolation. Error-detection benchmarks are t...
362. Turning spectra into images improves plant trait retrieval with 2D-CNNs ​
Author: Javier Lopatin, Teja Kattenborn, Eya Cherif, Sebasti'an Moreno
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.16661v1 Announce Type: cross Abstract: Hyperspectral reflectance spectroscopy enables non-destructive estimation of plant functional traits, yet current deep learning approaches process spectra as one-dimensional sequences, which limits how they capture long-range inter-band dependencies....
363. Random Quadratic Form with random forcing: Metastable synchronization by noise ​
Author: Anna Shalova
Published: 8/19/2026, 4:00:00 AM
Categories: math.PR, cs.LG, math.DS
arXiv:2608.16664v1 Announce Type: cross Abstract: We study the Random Quadratic Form (RQF) on a sphere in the presence of random Brownian forcing. We show that the forcing does not effectively change the law of the process but affects the synchronization properties of the system. While the RQF witho...
364. Hide&Seek: Learning to Explain in an End-to-End Differentiable Network ​
Author: Tal Ellinson, Hadi Mohasel Afshar, Sally Cripps
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.16689v1 Announce Type: cross Abstract: Instance-wise feature selection is a valuable tool for interpreting labeled data and the predictions of black-box models. In contrast to global feature selection techniques, instance-wise methods dynamically identify important features for each insta...
365. Learning to Price with Persuasion ​
Author: Maria-Florina Balcan, Tejas Pagare, Karan Singh
Published: 8/19/2026, 4:00:00 AM
Categories: cs.GT, cs.LG, econ.TH
arXiv:2608.16699v1 Announce Type: cross Abstract: Motivated by modern marketplaces, where the platform or the seller routinely gathers detailed user profiles, we study a novel learning theoretic model that simultaneously involves information and mechanism design. Specifically, we consider the econom...
366. MIRROR: Multimodal Intelligent Radiology Reasoning and Observation Reporter ​
Author: Vignesh Nagarajan, Sriram Venkatapathy
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.16709v1 Announce Type: cross Abstract: A radiologist reading a model's output faces two problems. The model returns a number and no reason, and any system that turns that number into readable prose can quietly add claims the model never made. MIRROR is a research prototype built to separa...
367. ClawGym II: Exploring Black-Box RL on Agent Harness ​
Author: Huatong Song, Fei Bai, Ming Yang, Renyuan Li, Jia Deng, Jujie He, Zhange Zhang, Daixuan Cheng, Yan Xing, Qi Yun, Xuxing Chen, Danyang Li, Feng Chang, Chuan Hao, Ran Tao, Jian Yang, Bryan Dai, Wayne Xin Zhao, Mingjie Tang, Ji-Rong Wen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.16798v1 Announce Type: cross Abstract: Agent harnesses have substantially improved performance on long-horizon tasks by coordinating agent interactions with the environment. However, reinforcement learning through complex harnesses remains largely unexplored, as scaling such training to l...
368. Unsupervised Learning of Cell Instances with Generative Routing Pyramids ​
Author: Ziwen Liu, Martin Weigert
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, q-bio.QM
arXiv:2608.16810v1 Announce Type: cross Abstract: Identifying and representing object instances such as cells or nuclei is a common task in microscopy image analysis. Established machine learning workflows typically use supervised detection or segmentation followed by feature extraction or classific...
369. zLend: A Dual-Scope Cash-Flow Reconstruction Framework for On-Chain Credit Underwriting ​
Author: Girish G N, Ashutosh Sahoo, Akshay SP, Gurukiran S, Dhanashekar Kandaswamy
Published: 8/19/2026, 4:00:00 AM
Categories: q-fin.RM, cs.LG
arXiv:2608.16856v1 Announce Type: cross Abstract: Decentralized lending lacks a credit bureau: a borrower's capacity to repay must be inferred entirely from public on-chain activity, without income verification or a liability record. This paper presents zLend, a deployed cash-flow underwriting frame...
370. The canonical facets of multi-separator polytopes ​
Author: Bjoern Andres, Silvia Di Gregorio, Jannik Irmai, Lucas Fabian Naumann, Shengxian Zhao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DM, cs.LG, math.CO
arXiv:2608.16861v1 Announce Type: cross Abstract: We initiate a polyhedral study of the graph multi-separator problem proposed by Irmai et al. (2024) as an alternative to the lifted multicut problem for application to the task of image segmentation. Starting with an integer linear program (ILP) form...
371. Non-Crossing Deep Quantile Regression for Distributional Survival Prediction ​
Author: Shuai Huang, Zhe Qu, Zhaowei Hua, Guohao Shen, Rui Tang, Hongtu Zhu
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP
arXiv:2608.16864v1 Announce Type: cross Abstract: In survival analysis the way covariates act on the risk of an event often differs between early and late failure times, yet hazard- and mean-based summaries collapse this variation into a single number. Quantile-based modeling instead describes the f...
372. AutoSR: Automatic Symbolic Regression by Searching Research States ​
Author: Kejia Zhang, Youran Sun, Xinyu Ren, Chugang Yi, Haizhao Yang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SC, cs.AI, cs.LG, cs.NA, math.NA
arXiv:2608.16876v1 Announce Type: cross Abstract: We introduce Automatic Symbolic Regression (AutoSR), a fully automated system that instantiates Research-Space Symbolic Regression by searching persistent scientific investigations rather than isolated equations. Finite, noisy data often yield numeri...
373. Spectral Gaps of Hit-and-Run and Coordinate Hit-and-Run ​
Author: Yunbum Kook, Santosh S. Vempala
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, math.PR, math.ST, stat.TH
arXiv:2608.16878v1 Announce Type: cross Abstract: For any convex body $\mathcal{K}\subset\mathbb{R}^{n}$ containing a unit ball, the spectral gap of Hit-and-Run is $\Omega(1/(n^2 C_{\mathsf{PI}}))$, where $C_{\mathsf{PI}}$ is the Poincar'e constant of the uniform distribution $\pi$ over $\mathcal{K...
374. Improving the matrix multiplication exponent with modern optimization and AlphaEvolve ​
Author: Emilien Dupont, Marvin Eisenberger, Borislav Kozlovskii, Abbas Mehrabian, Francisco J. R. Ruiz, Abigail See, Renfei Zhou, Josh Alman, Virginia Vassilevska Williams, Matej Balog
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DS, cs.AI, cs.CC, cs.LG
arXiv:2608.16884v1 Announce Type: cross Abstract: The current best bounds on the matrix multiplication exponent $\omega$ are obtained through a refinement of the laser method called combination loss analysis (Duan et al., 2022; Williams et al., 2024; Alman et al., 2025). In this note, we address the...
375. Sequential Batch Learning in Finite-Action Linear Contextual Bandits ​
Author: Yanjun Han, Zhengqing Zhou, Zihao Hu, Jose Blanchet, Peter W. Glynn, Yinyu Ye, Zhengyuan Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, stat.ML
arXiv:2004.06321v2 Announce Type: replace Abstract: We study the sequential batch learning problem in linear contextual bandits with finite action sets, where the decision maker is constrained to split incoming individuals into (at most) a fixed number of batches and can only observe outcomes for th...
376. Pointer Networks with Q-Learning for Combinatorial Optimization ​
Author: Alessandro Barro
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2311.02629v5 Announce Type: replace Abstract: We introduce the Pointer Q-Network (PQN), a hybrid neural architecture that integrates model-free Q-value policy approximation with Pointer Networks (Ptr-Nets) to enhance the optimality of attention-based sequence generation, focusing on long-term ...
377. DirMixE: Harnessing Test Agnostic Long-tail Recognition with Hierarchical Label Variations ​
Author: Zhiyong Yang, Qianqian Xu, Sicong Li, Zitai Wang, Xiaochun Cao, Qingming Huang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2405.07780v3 Announce Type: replace Abstract: This paper explores test-agnostic long-tail recognition, a challenging long-tail task where the test label distributions are unknown and arbitrarily imbalanced. We argue that the variation in these distributions can be broken down hierarchically in...
378. Balancing Optimality and Diversity: Human-Centered Decision Making through Generative Curation ​
Author: Michael Lingzhi Li, Shixiang Zhu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, math.OC
arXiv:2409.11535v3 Announce Type: replace Abstract: Many decision-support systems recommend actions by optimizing measurable objectives, even when a human decision-maker retains final authority and considers additional criteria that are difficult to specify in advance. We study how an algorithm shou...
379. Limits to scalable evaluation at the frontier: LLM as Judge won't beat twice the data ​
Author: Florian E. Dorner, Vivian Y. Nastl, Moritz Hardt
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2410.13341v4 Announce Type: replace Abstract: High quality annotations are increasingly a bottleneck in the explosively growing machine learning ecosystem. Scalable evaluation methods that avoid costly annotation have therefore become an important research ambition. Many hope to use strong exi...
380. MoE-Enhanced Explainable Deep Manifold Transformation for Complex Data Embedding and Visualization ​
Author: Zelin Zang, Yuhao Wang, Jinlin Wu, Hong Liu, Yue Shen, Zhen Lei, Stan Z. Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2410.19504v3 Announce Type: replace Abstract: Dimensionality reduction (DR) plays a crucial role in various fields, including data engineering and visualization, by simplifying complex datasets while retaining essential information. However, achieving both high DR accuracy and strong explainab...
381. Model Inversion Attacks: A Survey of Approaches and Countermeasures ​
Author: Zhanke Zhou, Jianing Zhu, Fengfei Yu, Xuan Li, Xiong Peng, Tongliang Liu, Bo Han
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2411.10023v3 Announce Type: replace Abstract: Deep neural networks have enabled numerous studies and applications on both Euclidean data, such as images and text, and non-Euclidean data, such as graphs. Because these networks may process private data, their deployment raises concerns about pri...
382. Rethinking Token-wise Feature Caching: Accelerating Diffusion Transformers with Dual Feature Caching ​
Author: Chang Zou, Shikang Zheng, Evelyn Zhang, Runlin Guo, Haohang Xu, Zhengyi Shi, Conghui He, Xuming Hu, Linfeng Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2412.18911v3 Announce Type: replace Abstract: Diffusion Transformers (DiT) have become the dominant methods in image and video generation yet still suffer substantial computational costs. As an effective approach for DiT acceleration, feature caching methods are designed to cache the features ...
383. ConfRetro: a 3D-aware template-free method for enhancing retrosynthesis via molecular conformer information ​
Author: Jiaxi Zhuang, Yu Zhang, Ying Qian, Aimin Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2501.12434v3 Announce Type: replace Abstract: Motivation: Retrosynthesis plays a crucial role in organic synthesis and drug discovery, focusing on identifying a set of reactants capable of synthesizing a target product molecule. Although the existing approaches have shown promising results, th...
384. Towards Unified Approaches in Self-Supervised Event Stream Modeling: Progress and Prospects ​
Author: Levente Z'olyomi, Tianze Wang, Sofiane Ennadir, Oleg Smirnov, Lele Cao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2502.04899v3 Announce Type: replace Abstract: The proliferation of digital interactions across diverse domains, such as healthcare, e-commerce, gaming, and finance, has resulted in the generation of vast volumes of event stream (ES) data. ES data comprises continuous sequences of timestamped e...
385. Leveraging Machine Unlearning for Cost-Efficient Preference Alignment ​
Author: Xiaohua Feng, Yuyuan Li, Huwei Ji, Jiaming Zhang, Li Zhang, Tianyu Du, Chaochao Chen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2504.06659v2 Announce Type: replace Abstract: Despite advances in Preference Alignment (PA) for Large Language Models (LLMs), mainstream methods like reinforcement learning with human feedback face notable challenges. These approaches require high-quality datasets of positive preference exampl...
386. A Generative Deep Learning Workflow for Inverse Molecular Design of Fuels ​
Author: Kiran K. Yalamanchi, Pinaki Pal, Balaji Mohan, Abdullah S. AlRamadan, Jihad A. Badra, Yuanjiang Pei
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph
arXiv:2504.12075v4 Announce Type: replace Abstract: In the present work, a generative deep learning framework combining a Co-optimized Variational Autoencoder (Co-VAE) with quantitative structure-property relationship (QSPR) techniques is developed to enable inverse molecular design of fuels. The Co...
387. Hybrid quantum recurrent neural network for remaining useful life prediction of turbofan engines ​
Author: Olga Tsurkan, Aleksandra Konstantinova, Arsenii Senokosov, Asel Sagingalieva, Alexey Melnikov
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, quant-ph
arXiv:2504.20823v3 Announce Type: replace Abstract: Accurate remaining useful life (RUL) estimation underpins safe operation and cost-effective maintenance of aerospace propulsion systems. We propose a Hybrid Quantum Recurrent Neural Network (HQRNN) for jet-engine RUL forecasting on the NASA C-MAPSS...
388. FDA-Opt: Federated Fine-Tuning via Dynamic Update Schedules ​
Author: Michael Theologitis, Vasilis Samoladas, Antonios Deligiannakis
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2505.04535v4 Announce Type: replace Abstract: Federated Learning (FL) enables the utilization of vast, previously inaccessible data sources. At the same time, pre-trained Language Models (LMs) have taken the world by storm and for good reason. They exhibit remarkable emergent abilities and are...
389. WATCH: Adaptive Monitoring for AI Deployments via Weighted-Conformal Martingales ​
Author: Drew Prinster, Xing Han, Anqi Liu, Suchi Saria
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2505.04608v5 Announce Type: replace Abstract: Responsibly deploying artificial intelligence (AI) / machine learning (ML) systems in high-stakes settings arguably requires not only proof of system reliability, but also continual, post-deployment monitoring to quickly detect and address any unsa...
390. One-shot Robust Federated Learning of Independent Component Analysis ​
Author: Dian Jin, Xin Bing, Yuqian Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML
arXiv:2505.20532v2 Announce Type: replace Abstract: This paper studies robust one-shot aggregation for distributed and federated Independent Component Analysis (ICA). In this setting, each client computes a local ICA estimator, while the server aims to recover a common global mixing matrix without a...
391. PhyxMamba: Chaotic System Reconstruction from Short Context Observations with Generative State-Space Models ​
Author: Chang Liu, Bohao Zhao, Jingtao Ding, Huandong Wang, Yong Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2505.23863v3 Announce Type: replace Abstract: Understanding chaotic dynamics is a fundamental problem across scientific disciplines, including climate science, neuroscience, and fluid dynamics, yet direct experimentation and intervention in such systems are often infeasible. Chaotic system rec...
392. Feature-Aware (Hyper)graph Generation via Next-Scale Prediction ​
Author: Dorian Gailhard, Enzo Tartaglione, Lirida Naviner, Jhony H. Giraldo
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.DM
arXiv:2506.01467v4 Announce Type: replace Abstract: Graph generative models perform well on small-scale structured data but struggle to scale to large, complex structures. Hierarchical approaches improve scalability but often ignore node and edge features, which are critical in real-world applicatio...
393. VirnyFlow: Optimizing ML Pipelines for Accuracy, Fairness, and Stability at Scale ​
Author: Denys Herasymuk, Anastasiia Mozghova, Nazar Protsiv, Vladyslav Sydorak, Julia Stoyanovich
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY
arXiv:2506.01584v2 Announce Type: replace Abstract: Developing machine learning (ML) systems for real-world deployment requires navigating context-dependent trade-offs among accuracy, fairness, stability, and other objectives. Existing AutoML frameworks optimize pipelines efficiently, but they fix t...
394. Contraction-Aware Reinforcement Learning for Nonlinear Control with Statistical Robustness ​
Author: Minjae Cho, Hiroyasu Tsukamoto, Huy T. Tran
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2506.15700v2 Announce Type: replace Abstract: Control contraction metrics (CCMs)-defined by Riemannian metrics under which a closed-loop system is incrementally exponentially stable-offer a constructive framework for synthesizing contracting policies in nonlinear path-tracking problems. Howeve...
395. Leveraging Generative Artificial Intelligence for Causal Inference with Unstructured Data ​
Author: Kosuke Imai, Kentaro Nakamura
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML
arXiv:2507.03897v4 Announce Type: replace Abstract: We introduce GenAI-Powered Inference (GPI), a statistical framework for both causal and predictive inference using unstructured data, including text and images. GPI leverages open-source Generative Artificial Intelligence (GenAI) models---such as l...
396. ROC-n-reroll: How verifier imperfection affects test-time scaling ​
Author: Florian E. Dorner, Yatong Chen, Andr'e F. Cruz, Fanny Yang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2507.12399v3 Announce Type: replace Abstract: Test-time scaling aims to improve language model performance by leveraging additional compute during inference. Many works have empirically studied techniques such as Best-of-N (BoN) and Rejection Sampling (RS) that make use of a verifier to enable...
397. Generative Model Unlearning: A Survey through Target Events, Unlearning Operators, and Evaluation Protocols ​
Author: Xiaohua Feng, Jiaming Zhang, Fengyuan Yu, Chengye Wang, Li Zhang, Kaixiang Li, Yuyuan Li, Lingjuan Lyu, Chaochao Chen, Jianwei Yin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2507.19894v2 Announce Type: replace Abstract: With the rapid advancement of generative models, privacy, copyright, safety, and reliability risks have attracted growing attention. To mitigate these risks, machine unlearning has been increasingly adapted from traditional classification models to...
398. Learning Multi-Timescale Interventions under Safety and Resource Constraints ​
Author: David Mguni, Wanrong Yang, Jing Dong, Jing Peng, Ziquan Liu, Muhammad Salman Haleem, Baoxiang Wang, Dominik Wojtczak
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2508.03875v4 Announce Type: replace Abstract: Many sequential decision problems offer qualitatively different ways of influencing the environment: some interventions act immediately, whereas others induce persistent effects that continue to shape future states long after the decision that init...
399. ProteoKnight: Convolution-based Phage Virion Protein Classification and Uncertainty Analysis ​
Author: Samiha Afaf Neha, Md. Ishrak Khan, Abir Ahammed Bhuiyan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2508.07345v2 Announce Type: replace Abstract: \textbf{Introduction:} Accurate prediction of Phage Virion Proteins (PVP) is essential for genomic studies due to their crucial role as structural elements in bacteriophages. Computational tools, particularly machine learning, have emerged for anno...
400. Enhancing Differentially Private Linear Regression via Public Second-Moment ​
Author: Zilong Cao (The School of Mathematics, Northwest University), Hai Zhang (The School of Mathematics, Northwest University)
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML
arXiv:2508.18037v2 Announce Type: replace Abstract: Leveraging information from public data has become increasingly crucial in enhancing the utility of differentially private (DP) methods. Traditional DP approaches often require adding noise based solely on private data, which can significantly degr...
401. Near-Optimal Sample Complexity Bounds for Constrained Average-Reward MDPs ​
Author: Yukuan Wei, Xudong Li, Lin F. Yang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2509.16586v2 Announce Type: replace Abstract: Recent advances have significantly improved our understanding of the sample complexity of learning in average-reward Markov decision processes (AMDPs) under the generative model. However, much less is known about the constrained average-reward MDP ...
402. Knowing When to Defer: Selective Prediction for Responsible Knowledge Tracing ​
Author: Joshua Mitton, Prarthana Bhattacharyya, Ralph Abboud, Simon Woodhead
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2509.21514v4 Announce Type: replace Abstract: Research on Knowledge Tracing (KT) models traditionally focuses on improving predictive accuracy. However, responsible real-world deployment requires models to know when to defer uncertain predictions to a human teacher. We introduce an intrinsic s...
403. Federated Self-Supervised Modulation Classification under Non-IID and Imbalanced Data ​
Author: Usman Akram, Yiyue Chen, Haris Vikalo
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP
arXiv:2510.04927v2 Announce Type: replace Abstract: Automatic modulation classification (AMC) is a core enabler of cognitive wireless systems, providing spectrum awareness and supporting adaptive communication at the network edge. However, training AMC models on centrally aggregated data incurs high...
404. Time-Correlated Video Bridge Matching ​
Author: Viacheslav Vasilev, Arseny Ivanov, Nikita Gushchin, Maria Kovaleva, Alexander Korotin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.12453v3 Announce Type: replace Abstract: Diffusion models excel in noise-to-data generation tasks, providing a mapping from a Gaussian distribution to a more complex data distribution. However, they struggle to model translations between complex distributions, limiting their effectiveness...
405. Near-Equilibrium Propagation training in nonlinear wave systems ​
Author: Karol Sajnok, Micha{\l} Matuszewski
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.quant-gas, math-ph, math.MP, physics.optics, quant-ph
arXiv:2510.16084v3 Announce Type: replace Abstract: Backpropagation learning algorithm, the workhorse of modern artificial intelligence, is notoriously difficult to implement in physical neural networks. Equilibrium Propagation (EP) is an alternative with comparable efficiency and strong potential f...
406. Explainable Heterogeneous Anomaly Detection in Financial Networks via Adaptive Expert Routing ​
Author: Zan Li, Rui Fan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE
arXiv:2510.17088v3 Announce Type: replace Abstract: Financial anomalies arise from heterogeneous mechanisms - price shocks, liquidity freezes, contagion cascades, and momentum reversals - yet existing detectors produce uniform anomaly scores without revealing which mechanism is failing or where risk...
407. Last-Iterate Analyses of FTRL with the 1/2-Tsallis Entropy in Stochastic Bandits ​
Author: Jingxin Zhan, Yuze Han, Zhihua Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.22819v3 Announce Type: replace Abstract: The convergence analysis of online learning algorithms is central to machine learning theory, where the last-iterate convergence is particularly important, as it captures the learner's actual decisions and describes the evolution of the learning pr...
408. Retrofit: Continual Learning with Controlled Forgetting for Binary Security Detection and Analysis ​
Author: Yiling He, Junchi Lei, Hongyu She, Shuo Shao, Xinran Zheng, Yiping Liu, Zhan Qin, Lorenzo Cavallaro
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2511.11439v3 Announce Type: replace Abstract: Binary security has increasingly relied on deep learning to reason about malware behavior and program semantics. However, the performance often degrades as threat landscapes evolve and code representations shift. While continual learning (CL) offer...
409. Q-Regularized Generative Auto-Bidding: From Suboptimal Trajectories to Optimal Policies ​
Author: Mingming Zhang, Na Li, Zhuang Feiqing, Hongyang Zheng, Jiangbing Zhou, Wang Wuyin, Sheng-jie Sun, XiaoWei Chen, Junxiong Zhu, Lixin Zou, Chenliang Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR
arXiv:2601.02754v3 Announce Type: replace Abstract: With the rapid development of e-commerce, auto-bidding has become a key asset in optimizing advertising performance under diverse advertiser environments. The current approaches focus on reinforcement learning (RL) and generative models. These effo...
410. Replacing Tunable Parameters in Weather and Climate Models with State-Dependent Functions using Reinforcement Learning ​
Author: Pritthijit Nath, Sebastian Schemm, Henry Moss, Peter Haynes, Emily Shuckburgh, Mark J. Webb
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph
arXiv:2601.04268v4 Announce Type: replace Abstract: Weather and climate models rely on parametrisations to represent unresolved sub-grid processes. Traditional schemes rely on fixed coefficients that are weakly constrained and tuned offline, contributing to persistent biases that limit their ability...
411. Stitch the Fragments: One-Shot Hierarchical Federated Clustering ​
Author: Shenghong Cai, Zihua Yang, Yang Lu, Mengke Li, Yuzhu Ji, Yiqun Zhang, Yiu-Ming Cheung
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.06404v2 Announce Type: replace Abstract: Federated Clustering (FC) faces a critical bottleneck in real-world scenarios, i.e., global clusters are rarely intact, often fragmenting into incomplete, multi-granular unlabeled ``clusterlets'' distributed across Non-IID clients. Although hierarc...
412. In-Context Source and Channel Coding ​
Author: Ziqiong Wang, Tianqi Ren, Rongpeng Li, Zhifeng Zhao, Honggang Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.10267v2 Announce Type: replace Abstract: Separate Source-Channel Coding (SSCC) remains attractive for text transmission due to its modularity and compatibility with mature entropy coders and powerful channel codes. However, SSCC often suffers from a pronounced cliff effect in low Signal-t...
413. Robust Privacy: Inference-Stage Privacy through Certified Robustness ​
Author: Jiankai Jin, Xiangzheng Zhang, Zhao Liu, Wenzhuo Xu, Dongdong Yang, Deyue Zhang, Quanchen Zou
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2601.17360v3 Announce Type: replace Abstract: An adversary observing a model's released prediction can infer sensitive attributes of the queried input, or even reconstruct representatives of the model's training data. The inference interface thus acts as a side channel for privacy leakage. We ...
414. Analytical Provisioning for Attention-FFN Disaggregated LLM Serving under Stochastic Workloads ​
Author: Chendong Song, Meixuan Wang, Hang Zhou, Hong Liang, Yuan Lyu, Zixi Chen, Yuwei Fan, Zijie Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2601.21351v4 Announce Type: replace Abstract: Attentio-FFN disaggregation (AFD) is an emerging architecture for LLM decoding that separates state-heavy, KV-cache-dominated Attention computation from stateless, compute-intensive FFN computation, connected by per-step communication. While AFD en...
415. Evaluating the trustworthiness of the Fr\'echet Inception Distance with stochastic embedding representations ​
Author: Ciaran Bench, Vivek Desai, Carlijn Roozemond, Ruben van Engen, Spencer A. Thomas
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.21979v2 Announce Type: replace Abstract: Feature embeddings acquired from pretrained models are widely used in medical applications of deep learning to assess the characteristics of datasets; e.g. to determine the quality of synthetic, generated medical images. The Fr'{e}chet Inception D...
416. PRO-Bid: Pareto-Prioritized Regret Optimization for Constraint-Aware Generative Auto-Bidding ​
Author: Binglin Wu, Yingyi Zhang, Xianneng Li, Ruyue Deng, Chuan Yue, Weiru Zhang, Xiaoyi Zeng
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.GT
arXiv:2602.08261v2 Announce Type: replace Abstract: Auto-bidding systems strive to maximize marketing value while maintaining high compliance with efficiency constraints, such as Target Cost-Per-Action (CPA). While Decision Transformers offer powerful sequence modeling capabilities, their applicatio...
417. Grounding LTL Tasks in Sub-Symbolic RL Environments for Zero-Shot Generalization ​
Author: Matteo Pannacci, Andrea Fanti, Elena Umili, Roberto Capobianco
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.09761v2 Announce Type: replace Abstract: In this work we address the problem of training a Reinforcement Learning agent to follow multiple temporally-extended instructions expressed in Linear Temporal Logic in sub-symbolic environments. Previous multi-task work has mostly relied on knowle...
418. Adaptive Optimization via Momentum on Variance-Normalized Gradients ​
Author: Francisco Patitucci, Aryan Mokhtari
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2602.10204v2 Announce Type: replace Abstract: We introduce MVN-Grad (Momentum on Variance-Normalized Gradients), an Adam-style optimizer that improves stability and performance by combining two complementary ideas: variance-based normalization and momentum applied after normalization. MVN-Grad...
419. Closing the Loop: A Control-Theoretic Framework for Provably Stable Time Series Forecasting with LLMs ​
Author: Xingyu Zhang, Jingyao Wang, Zeen Song, Changwen Zheng, Wenwen Qiang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.12756v2 Announce Type: replace Abstract: Large Language Models (LLMs) have recently shown exceptional potential in time series forecasting (TSF), leveraging their inherent sequential reasoning capabilities to model complex temporal dynamics. Existing approaches typically employ an autoreg...
420. Zero-Shot Instruction Following in RL via Structured LTL Representations ​
Author: Mathias Jackermeier, Mattia Giuri, Jacques Cloete, Alessandro Abate
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.14344v2 Announce Type: replace Abstract: We study instruction following in multi-task reinforcement learning, where an agent must zero-shot execute novel tasks not seen during training. In this setting, linear temporal logic (LTL) has recently been adopted as a powerful framework for spec...
421. MDP Planning as Policy Inference ​
Author: David Tolpin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.17375v3 Announce Type: replace Abstract: We formulate episodic Markov decision process (MDP) planning as Bayesian inference over policies. The primary contribution is conceptual: the policy itself is treated as the latent variable, and expected return defines an unnormalized posterior den...
422. LORA-CRAFT: Cross-layer Rank Adaptation via Frozen Tucker Decomposition of Pre-trained Attention Weights ​
Author: Kasun Dewage, Marianna Pensky, Suranadi De Silva, Shankadeep Mondal
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.17510v2 Announce Type: replace Abstract: We introduce LoRA-CRAFT (\textbf{C}ross-layer \textbf{R}ank \textbf{A}daptation via \textbf{F}rozen \textbf{T}ucker), abbreviated CRAFT throughout, an extremely parameter-efficient fine-tuning (PEFT) method that applies Tucker tensor decomposition ...
423. Measuring the Prevalence of Policy Violating Content with ML Assisted Sampling and LLM Labeling ​
Author: Attila Dobi, Aravindh Manickavasagam, Benjamin Thompson, Xiaohan Yang, Faisal Farooq
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML
arXiv:2602.18518v2 Announce Type: replace Abstract: Content safety teams need metrics that reflect what users actually experience, not only what is reported. We study prevalence: the fraction of user views (impressions) that went to content violating a given policy on a given day. Accurate prevalenc...
424. Exact Attention Sensitivity and the Geometry of Transformer Stability ​
Author: Seyed Morteza Emadi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.18849v2 Announce Type: replace Abstract: We develop a sensitivity analysis for transformer attention in a geometry aligned with tokenwise computation. Our main result is the exact identity $|J_\tau(u)|{\infty\to1}=\theta(p)/\tau$ for the Jacobian $J\tau(u)$ of the tempered softmax $u...
425. Decentralized Federated Learning by Partial Message Exchange ​
Author: Shan Sha, Shenglong Zhou, Xin Wang, Lingchen Kong, Geoffrey Ye Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.01730v4 Announce Type: replace Abstract: Decentralized federated learning (DFL) has emerged as a transformative server-free paradigm that enables collaborative learning over large-scale heterogeneous networks. However, it continues to face fundamental challenges, including data heterogene...
426. Joint MDPs and Reinforcement Learning in Coupled-Dynamics Environments ​
Author: Ege C. Kaya, Mahsa Ghasemi, Abolfazl Hashemi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2603.06946v2 Announce Type: replace Abstract: Many distributional quantities in reinforcement learning are intrinsically joint across actions, including distributions of gaps and probabilities of superiority. However, the classical Markov decision process (MDP) formalism specifies only margina...
427. Not All Neighbors Matter: Understanding the Impact of Graph Sparsification on GNN Pipelines ​
Author: Yuhang Song, Naima Abrar Shami, Romaric Duvignau, Vasiliki Kalavri
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.DB
arXiv:2603.06952v2 Announce Type: replace Abstract: As graphs scale to billions of nodes and edges, graph Machine Learning workloads are constrained by the cost of multi-hop traversals over exponentially growing neighborhoods. While various system-level and algorithmic optimizations have been propos...
428. Differentiable Thermodynamic Phase-Equilibria for Machine Learning ​
Author: Karim K. Ben Hicham, Moreno Ascani, Jan G. Rittig, Alexander Mitsos
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.11249v4 Announce Type: replace Abstract: Accurate prediction of phase equilibria remains a central challenge in chemical engineering. Physics-consistent machine learning methods that incorporate thermodynamic structure into neural networks have recently shown strong performance for activi...
429. Informative Perturbation Selection for Uncertainty-Aware Post-hoc Explanations ​
Author: Sumedha Chugh, Ranjitha Prasad, Nazreen Shah
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2603.14894v3 Announce Type: replace Abstract: Trust and ethical concerns due to the widespread deployment of opaque machine learning (ML) models motivating the need for reliable model explanations. Post-hoc model-agnostic explanation methods addresses this challenge by learning a surrogate mod...
430. Free-Flow Class-Incremental Learning: Towards Robust CIL under Variable Class Arrivals ​
Author: Zhiming Xu, Baile Xu, Jian Zhao, Furao Shen, Suorong Yang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.02765v2 Announce Type: replace Abstract: Class-incremental learning (CIL) is commonly evaluated under predefined schedules with fixed or nearly equal class increments, leaving irregular class-arrival scenarios underexplored. However, practical CIL systems may need to update whenever new c...
431. A Theoretical Framework for Statistical Evaluability of Generative Models ​
Author: Shashaank Aiyer, Yishay Mansour, Shay Moran, Han Shao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2604.05324v3 Announce Type: replace Abstract: Statistical evaluation aims to estimate the generalization performance of a model using held-out i.i.d. test data sampled from the ground-truth distribution. In supervised learning settings such as classification, performance metrics such as error ...
432. From Uniform to Learned Knots: A Study of Spline-Based Numerical Encodings for Tabular Deep Learning ​
Author: Manish Kumar, Anton Frederik Thielmann, Christoph Weisser, Benjamin S"afken
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.05635v2 Announce Type: replace Abstract: Numerical preprocessing remains a critical component of tabular deep learning, as the representation of continuous features can strongly affect downstream performance. We systematically study spline-based numerical encodings, including B-splines, M...
433. CodeQuant: Unified Clustering and Quantization for Enhanced Outlier Smoothing in Low-Precision Mixture-of-Experts ​
Author: Xiangyang Yin, Xingyu Liu, Tianhua Xia, Bo Bao, Vithursan Thangarasa, Valavan Manohararajah, Eric Sather, Sai Qian Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.10496v2 Announce Type: replace Abstract: Outliers have emerged as a fundamental bottleneck in preserving accuracy for low-precision large models, particularly within Mixture-of-Experts (MoE) architectures that are increasingly central to large-scale language modeling. Under post-training ...
434. The Optimal Sample Complexity of Multiclass and List Learning ​
Author: Chirag Pabbaraju
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2604.24749v4 Announce Type: replace Abstract: While the optimal sample complexity of binary classification in terms of the VC dimension is well-established, determining the optimal sample complexity of multiclass classification has remained open. The appropriate complexity parameter for multic...
435. Learning Koopman operators for coupled systems via information on governing equations of subsystems ​
Author: Tatsuya Naoi, Jun Ohkubo
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.01835v2 Announce Type: replace Abstract: Nonlinear coupled systems are ubiquitous in science and engineering. The analysis and modeling of such systems are challenging due to their high dimensionality and complex interactions among subsystems. In recent years, operator-theoretic methods b...
436. Delay, Plateau, or Collapse: Evaluating the Impact of Systematic Verification Error on RLVR ​
Author: Kazuki Egashira, Mark Vero, Jasper Dekoninck, Florian E. Dorner, Robin Staab, Martin Vechev
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.02909v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a powerful approach for improving the reasoning capabilities of large language models (LLMs). While RLVR is designed for tasks with verifiable ground-truth answers, real-world verifie...
437. Can Attribution Predict Risk? From Multi-View Attribution to Planning Risk Signals in End-to-End Autonomous Driving ​
Author: Le Yang, Haijun Liu, Jiawei Liang, ShangQuan Sun, Xiaochun Cao
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.06264v2 Announce Type: replace Abstract: End-to-end autonomous driving models generate future trajectories from multi-view inputs, improving system integration but introducing opaque decisions and hard-to-localize risks. Existing methods either rely on auxiliary monitoring models or gener...
438. Convex Optimization with Nested Evolving Feasible Sets ​
Author: Karthick Krishna M., Haricharan Balasundaram, Rahul Vaze
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, math.OC
arXiv:2605.07386v2 Announce Type: replace Abstract: \emph{Convex Optimization with Nested Evolving Feasible Sets (CONES)} is considered where the objective function (f) remains fixed but the feasible region evolves over time as a nested sequence (S_1 \supseteq S_2 \supseteq \cdots \supseteq S_T)...
439. Learning to Bid with Unknown Private Values in Budget-Constrained First-Price Auctions ​
Author: Zihao Hu, Yuxiao Wen, Yuan Yao, Jiheng Zhang, Zhengyuan Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.09448v2 Announce Type: replace Abstract: We study the operational problem of automated bidding in repeated first-price auctions under budget and return-on-spend (RoS) constraints. In this setting, an auto-bidder must translate advertiser goals and constraints into real-time bids while lea...
440. Scaling Laws for Dynamic Mini-Batch SGD in Sketched Linear Regression ​
Author: Ziyan Chen, Zhongzhu Zhou, Ding-Xuan Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.24316v3 Announce Type: replace Abstract: Mini-batching is central to large-scale optimization, yet its role in statistical scaling laws remains limited. We study one-pass and multi-pass batch SGD for sketched linear regression under power-law spectral and source conditions. Our analysis r...
441. RT-Lynx: Putting GEMM Sparsity in the Right Place for Diffusion Models ​
Author: Xing Cong, Hanlin Tang, Kan Liu, Tao Lan, Lin Qu, Chenhao Xie
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.26632v3 Announce Type: replace Abstract: Diffusion Transformers (DiT) achieve strong performance in image generation but incur substantial inference costs. While prior work has reduced this cost via quantization and distillation, semi-structured sparsity, which can nearly halve FLOPs, rem...
442. Periodic Topological Deep Learning for Polymer Design and Discovery ​
Author: Yasharth Yadav, Tze Kwang Gerald Er, Atsushi Goto, Kelin Xia
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.26833v2 Announce Type: replace Abstract: Polymers underpin applications across energy, healthcare, and materials science, yet their vast chemical space makes systematic discovery challenging. Most machine learning approaches represent polymers as molecular graphs of a single repeating uni...
443. The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level Energy Attribution ​
Author: Deepak Panigrahy, Aakash Tyagi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR, cs.DC, cs.PF
arXiv:2605.27599v3 Announce Type: replace Abstract: Agentic AI workloads - where a single user goal triggers multi-step orchestration, tool calls, retries, and failure recovery - are being targeted for edge deployment, with NVIDIA, Dell, HP, ASUS, MSI, Acer, and Gigabyte all shipping GB10-based desk...
444. Representation Signatures and Risk-Feedback Alignment in LLM Trading Agents ​
Author: Weicheng Xue
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP
arXiv:2605.28850v3 Announce Type: replace Abstract: We study behavioral alignment and representation dynamics of large language model (LLM) agents in financial decision environments. TradeArena, an auditable trading-agent testbed with risk reports, execution simulation, memory, and replayable trajec...
445. Annealed Softmax Greedy in Many-Armed Bayesian Bandits ​
Author: William Overman, Mohsen Bayati
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.31034v3 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards and group-based policy optimization methods update a stochastic policy by sampling multiple completions per prompt and increasing the policy's probability on those with higher reward. These updates, un...
446. Learning symplectic model reduction based on an approximation theorem of symplectic embeddings ​
Author: Liyi Feng, Yifa Tang, Yulin Xie, Ruili Zhang, Aiqing Zhu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.04623v2 Announce Type: replace Abstract: High-dimensional Hamiltonian systems play a central role in many scientific and engineering disciplines, with dynamics that evolve on symplectic manifolds. Although deep learning provides powerful tools for constructing its low-dimensional surrogat...
447. Provably Efficient Personalized Multi-Objective Bandits with Proactive Conversational Queries ​
Author: Linfeng Cao, Ming Shi, Ness B. Shroff
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.08410v2 Announce Type: replace Abstract: Personalized decision-making in multi-objective bandits requires learning user-specific trade-offs among competing objectives. Since arm utility depends on both unknown rewards and unknown preferences, existing methods infer preferences only from u...
448. Speculative Rollback Correction for Quality-Diverse Web Agent Imitation ​
Author: Longkun Hao, Hongyu Lin, Hao Li, Zhuowen Liu, Zhichao Yang, Haojie Hao, Dongshuo Huang, Haitao Yang, Hongyu Ge, Ming jie Xie, Yanjun Wu, Zi Hao Yin, Yan Bai, Yihang Lou
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.12485v2 Announce Type: replace Abstract: Training interactive web agents through imitation learning from expert trajectories has emerged as a highly effective approach. However, determining the optimal timing for expert intervention presents a critical challenge in this context. Delayed i...
449. Sum-of-Squares Degree Barriers for the Reweighted-Hinge Method in Robust Halfspace Learning: A Christoffel-Function Characterization ​
Author: Xiaoyu Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, stat.ML
arXiv:2606.17215v2 Announce Type: replace Abstract: A certificate that removes outliers sees the data only through its low-degree moments, and an adversary exploits exactly this, hiding corruption where the clean data already looks typical, in the blind spot no bounded-degree test resolves. That bli...
450. Spectral Certificates and Projection-DPP Rounding for Determinantal MAP Selection ​
Author: Richard Yi Da Xu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.19411v3 Announce Type: replace Abstract: Selecting a fixed-size subset that maximizes the determinant of a positive semidefinite kernel is the MAP problem for a size-constrained determinantal point process and the classical maximum-entropy sampling problem. Although this discrete problem ...
451. SL-S4Wave: Self-Supervised Learning of Physiological Waveforms with Structured State Space Models ​
Author: Feng Wu, Harsh Deep, Eric Lehman, Sanyam Kapoor, Guoshuai Zhao, Rahul G. Krishnan, Gari Clifford, Li-wei H Lehman
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.19888v2 Announce Type: replace Abstract: Modeling long-sequence medical time series data, such as electrocardiograms (ECG), poses significant challenges due to high sampling rates, multichannel signal complexity, inherent noise, and limited labeled data. While recent self-supervised learn...
452. LACE-SVD: Loss-Aware SVD with Cumulative Error Correction for LLM Compression ​
Author: Zhuowen Liu, Longkun Hao, Shiyu Feng, Xiaowen Chang, Ruiqun Li, Changqun Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.03057v2 Announce Type: replace Abstract: The rapid growth in the parameter scale of large language models (LLMs) has created a strong demand for efficient compression techniques. As a hardware-agnostic and highly compatible approach, low-rank compression has been widely adopted to reduce ...
453. Directional Curvature from Armijo Backtracking: A Low-Cost Sharpness Probe and a Calibration-Free Learning-Rate Safeguard for Adam ​
Author: Ashmitha R, J"org Frochte
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.03998v4 Announce Type: replace Abstract: The local sharpness of the loss, the top Hessian eigenvalue $\lambda_1$, determines the largest stable gradient step, but measuring it normally requires Lanczos or Hessian-vector products. A single Armijo backtracking line search already carries th...
454. Efficient Safety Alignment of Language Models via Latent Personality Traits ​
Author: Mohamed Amine Merzouk, Nolan Smyth, Damiano Fornasiere, Linh Le, David Williams-King, Adam Oberman
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CR
arXiv:2607.07918v2 Announce Type: replace Abstract: Current safety methods for large language models are known to be vulnerable to adversarial attacks, motivating research into robust alternatives. Latent Adversarial Training (LAT) is among the most effective defenses, but can degrade utility and re...
455. LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning ​
Author: Ning Liu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.10139v3 Announce Type: replace Abstract: Selecting the correct answer from a pool of candidate reasoning chains is the engine of test-time scaling, yet the standard selectors each carry a cost: self-consistency inherits the errors of the single model it resamples, and trained reward model...
456. Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values ​
Author: Jan Betley, Johannes Treutlein, Jan Dubi'nski, Harry Mayne, Karol Ga{\l}\k{a}zka, Niels Warncke, Anna Sztyber-Betley, Owain Evans
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2607.14345v4 Announce Type: replace Abstract: People use language models for practical questions whose answers are difficult to verify. We show that models exhibit covert value leakage: the information they provide is influenced by their own values, without this influence being disclosed to th...
457. A Blueprint for Equilibrium-Based Differentiable Continuous-Variable Thermodynamic Computing ​
Author: Owen Lockwood, J'er'emy B'ejanin, Joost Bus, Christopher Chamberland, Patrick Huembeli, Frank Sch"afer, Guillaume Verdon
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.ET, physics.app-ph
arXiv:2607.16183v2 Announce Type: replace Abstract: To help address the escalating energy and latency demands of machine-learning workloads, we introduce a blueprint for an energy-efficient and fast thermodynamic computing stack that leverages stochastic analog processes in physical hardware. In thi...
458. SCPP: A Unified Python Library for Soft Clustering ​
Author: Kiyan Rezaee, Morteza Ziabakhsh, Artin Bahrampour, Seyed Mohammad Ghoreishi, Asal Khaje, Ali Sajedifar, Manny Chalak, Ava Zerafatangiz, Sadegh Eskandari
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.19620v2 Announce Type: replace Abstract: In this paper, we present SCPP (Soft Clustering Python Package), an open-source Python framework for soft clustering. SCPP establishes a canonical, scikit-learn-compatible estimator interface that standardizes model training, prediction, membership...
459. Global Convergence of DGM and PINN Algorithms for Solving Nonlinear PDEs ​
Author: Justin Sirignano, Konstantinos Spiliopoulos, Samuel Cohen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2607.24726v2 Announce Type: replace Abstract: The Deep Galerkin Method (DGM) and Physics Informed Neural Networks (PINNs) have become widely-used methods for solving partial differential equations (PDEs) in the rapidly growing field of scientific machine learning. In these methods, a neural ne...
460. Neurai-VN Benchmark: Standardized Machine Learning Models for Multimodal Digital Phenotyping in Mental Health Classification ​
Author: Quoc-Cuong Pham, Hoang-Thuy-Duong Vu, Thi-Thanh-Huong Ha, Huy-Hieu Pham
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.25232v3 Announce Type: replace Abstract: Digital phenotyping (DP) using smartphones and wearable devices has emerged as a promising approach for assessing mental health, particularly depression and anxiety. However, progress remains difficult to evaluate because of heterogeneity across da...
461. ODEWorld: A Continuous Predictive Architecture via Physical-Time Flow ​
Author: Dongxiu Liu, Haoyi Niu, Peng Cheng, Yuan Gao, Xirui Kang, Sangli Teng, Koushil Sreenath, Xianyuan Zhan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.RO
arXiv:2607.27924v3 Announce Type: replace Abstract: In the physical world we inhabit, space and time are fundamentally continuous. However, existing machine learning paradigms for world modeling are largely confined to discrete-time prediction, thereby exhibiting significant inefficiency in capturin...
462. Hierarchical Copula-Gumbel-Top-K Routing: Two-Sided Dependence Control for Frozen Mixture-of-Experts at Fixed Per-Token Routing Laws ​
Author: Richard Yi Da Xu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.28670v3 Announce Type: replace Abstract: A stochastic Gumbel-Top-K router defines, for every token of a mixture-of-experts (MoE) model, a routing law: a distribution over ordered expert lists and mixture weights. We ask which joint distributions over the routing choices of different token...
463. Technological Advances in Detecting and Managing Cognitive Impairment in Older Adults: Trends, Challenges, and Future Directions ​
Author: Mohammad Asif, Azizuddin Khan, Mohd Azam, Anurag Rajkumar Bombarde
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET
arXiv:2607.28687v2 Announce Type: replace Abstract: As populations age, cognitive decline from mild cognitive impairment (MCI) to dementia is a defining health challenge of the coming decades, yet routine assessment often misses its earliest signs. This article critically synthesizes recent technolo...
464. Learning Optimal Dynamic Matching via Graph Neural Networks ​
Author: Genta Okada, Shunya Noda, Junpei Komiyama, Akira Matsushita
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.GT, econ.TH
arXiv:2607.28925v2 Announce Type: replace Abstract: Dynamic matching markets require decisions about whom to match and when: matching now yields value but removes participants who may create better future opportunities. We develop a value-based reinforcement-learning framework for this problem on fi...
465. A 2-Block Architecture for Real-Time EEG Gait Decoding: A Pilot Study ​
Author: Shantanu Sarkar, Saurabh Prasad, Jose L. Contreras-Vidal
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, eess.SP
arXiv:2608.02083v2 Announce Type: replace Abstract: Closed-loop lower-limb exoskeleton control via Electroencephalography (EEG) remains limited by motion artifacts, low signal-to-noise ratio, and binary gait formulations that fail to capture full cortical gait complexity. We propose a 2-block Brain-...
466. Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't ​
Author: Ravi Satya Durga Prasad Yenugula
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.02829v3 Announce Type: replace Abstract: Model families are typically trained size by size, each from scratch. Can a pretrained large model instead be converted into a smaller sibling? We characterize the 1.4B->410M conversion in Pythia end to end. Representations align strongly across si...
467. A Physics-Informed Hybrid Neural Operator for Transient Magnetization Prediction in Power Magnetics ​
Author: Yachao Zhu, Qiujie Huang, Sinan Li, Yang Li, Gang Lei, Jianguo Zhu
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.SY, eess.SY
arXiv:2608.02965v3 Announce Type: replace Abstract: Magnetic components in high-frequency, high-power-density converters are increasingly driven by non-sinusoidal flux-density waveforms with fast transitions, minor-loop operation, dc bias, and temperature variation. Under these conditions, steady-st...
468. SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation ​
Author: Zikun Qu, Min Zhang, Mingze Kong, Zhiwei Shang, Zhengyu Chen, Yikun Ban, Shuang Qiu, Zhongxiang Dai
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.04419v2 Announce Type: replace Abstract: On-policy distillation (OPD) provides dense teacher supervision on student-generated trajectories, but standard reverse-KL training can assign insufficient probability to other plausible continuations. Teacher entropy alone does not reveal whether ...
469. A Probabilistic Circuit-Induced Pseudo-Metric for Out-of-Distribution Detection ​
Author: Bhumika K, Vidhya S, Narayanan C Krishnan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.09117v2 Announce Type: replace Abstract: Probabilistic Circuits (PCs) are tractable generative models whose internal nodes encode a hierarchy of probabilistic summaries over different variable scopes. Existing PC-based out-of-distribution (OOD) detection methods ignore this hierarchy, red...
470. From Recoverability to Functional Use: Auditing Temporal Reports in Time-Series Forecasting ​
Author: Qipeng Qian, Yuntao Qian
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.10433v4 Announce Type: replace Abstract: Time-series forecasters increasingly accompany numerical predictions with explicit temporal reports, such as delays or selected history, but a correct report need not describe the information actually used by the forecast. We separate this problem ...
471. JAPE: Joint Anomaly Prediction and Intrinsic Explanation in Multivariate Time Series ​
Author: Yian Wei, Yuanyuan Yao, Lu Chen, Xiangmin Zhou, Tianyi Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11801v2 Announce Type: replace Abstract: Multivariate time-series anomaly prediction aims to identify whether and when anomalies will occur over a future horizon from historical observations. Existing methods primarily characterize anomalies as deviations in future numerical values, which...
472. Represent, Then Generate: Multimodal-Conditioned Time-Series Generation under Irregular Missingness ​
Author: Haochen Zhang, Jiaheng Guo, Yu-Chao Huang, Nicholas Konz, Tianlong Chen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12592v2 Announce Type: replace Abstract: Continuous physiological time series underpin modern clinical monitoring, yet many of the most informative signals are invasive, expensive, or simply unavailable for a given patient. Conditional generation offers a remedy: an absent signal can be s...
473. From Fixed Grids to Moving Particles:A Transferable Latent Operator for Fluid Dynamics ​
Author: Meng Li, Chuqi Chen, Zhengqing Gao, Xi Zhou, Xiao Sun, Yang Xiang, Huaxi Huang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GR
arXiv:2608.14120v2 Announce Type: replace Abstract: Lagrangian modeling is vital to fluid dynamics, as it characterizes particle transport and complements the Eulerian representation. However, Lagrangian trajectories are less commonly available than Eulerian fields, while most neural operators are t...
474. From Monte Carlo to neural networks approximations of boundary value problems ​
Author: Lucian Beznea, Iulian Cimpean, Oana Lupascu-Stamate, Ionel Popescu, Arghir Zarnescu
Published: 8/19/2026, 4:00:00 AM
Categories: math.PR, cs.AI, cs.LG, cs.NA, math.AP, math.NA
arXiv:2209.01432v4 Announce Type: replace-cross Abstract: In this paper we study probabilistic and neural network approximations for solutions to Poisson equation subject to Holder data in general bounded domains of $\mathbb{R}^d$. We aim at two fundamental goals. The first, and the most important, ...
475. Doubly robust nearest neighbors in factor models ​
Author: Raaz Dwivedi, Caleb Chin, Sabina Tomkins, Predrag Klasnja, Susan Murphy, Devavrat Shah
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2211.14297v5 Announce Type: replace-cross Abstract: We introduce and analyze an improved variant of nearest neighbors (NN) for estimation with missing data in latent factor models. We consider a matrix completion problem with missing data, where the $(i, t)$-th entry, when observed, is given b...
476. EquiPocket: an E(3)-Equivariant Geometric Graph Neural Network for Ligand Binding Site Prediction ​
Author: Yang Zhang, Zhewei Wei, Ye Yuan, Chongxuan Li, Wenbing Huang
Published: 8/19/2026, 4:00:00 AM
Categories: q-bio.BM, cs.LG
arXiv:2302.12177v5 Announce Type: replace-cross Abstract: Predicting the binding sites of target proteins plays a fundamental role in drug discovery. Most existing deep-learning methods consider a protein as a 3D image by spatially clustering its atoms into voxels and then feed the voxelized protein...
477. Score Attack: A Lower Bound Technique for Optimal Differentially Private Learning ​
Author: T. Tony Cai, Yichen Wang, Linjun Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: math.ST, cs.CR, cs.LG, stat.ME, stat.ML, stat.TH
arXiv:2303.07152v3 Announce Type: replace-cross Abstract: Achieving optimal statistical performance while ensuring the privacy of personal data is a challenging yet crucial objective in modern data analysis. However, characterizing the optimality, particularly the minimax lower bound, under privacy ...
478. Dynamic Pricing and Advertising with Demand Learning ​
Author: Shipra Agrawal, Yiding Feng, Wei Tang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2304.14385v4 Announce Type: replace-cross Abstract: We consider a novel pricing and advertising framework in which a seller not only sets the product price but also designs flexible advertising schemes to influence customers' valuations of the product. We impose no structural restriction on th...
479. ODTlearn: A Package for Learning Optimal Decision Trees for Prediction and Prescription ​
Author: Patrick Vossler, Nathan Justin, Sina Aghaei, Nathanael Jo, Andr'es G'omez, Phebe Vayanos
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC
arXiv:2307.15691v4 Announce Type: replace-cross Abstract: ODTlearn is an open source Python package that provides methods for learning optimal decision trees for high-stakes predictive and prescriptive tasks based on the state-of-the-art mixed-integer optimization (MIO) framework proposed in Aghaei ...
480. ShadowNet for Data-Centric Quantum System Learning ​
Author: Yuxuan Du, Yibo Yang, Tongliang Liu, Zhouchen Lin, Bernard Ghanem, Dacheng Tao
Published: 8/19/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG
arXiv:2308.11290v2 Announce Type: replace-cross Abstract: Understanding the dynamics of large quantum systems is hindered by the curse of dimensionality. Statistical learning offers new possibilities in this regime through neural network protocols and classical shadows, while both methods have limit...
481. Macroeconomic Forecasting with Large Language Models ​
Author: Andrea Carriero, Davide Pettenuzzo, Shubhranshu Shekhar
Published: 8/19/2026, 4:00:00 AM
Categories: econ.EM, cs.CL, cs.LG
arXiv:2407.00890v5 Announce Type: replace-cross Abstract: This paper presents a comparative analysis evaluating the accuracy of Large Language Models (LLMs) against traditional macro time series forecasting approaches. In recent times, LLMs have surged in popularity for forecasting due to their abil...
482. Quantum Large Language Models via Tensor Network Disentanglers ​
Author: Borja Aizpurua, Fernando Loren, Saeed S. Jahromi, Sukhbinder Singh, Roman Orus
Published: 8/19/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG
arXiv:2410.17397v2 Announce Type: replace-cross Abstract: We introduce a framework for seamlessly integrating quantum computing into pretrained large language models (LLMs). The key idea is to construct a hybrid quantum-classical representation that exactly reproduces the original model, providing a...
483. Multi-Bin Batching for Increasing LLM Inference Throughput ​
Author: Ozgur Guldogan, Jackson Kunde, Kangwook Lee, Ramtin Pedarsani
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.DC, cs.LG, cs.SY, eess.SY
arXiv:2412.04504v2 Announce Type: replace-cross Abstract: As large language models (LLMs) grow in popularity for their diverse capabilities, improving the efficiency of their inference systems has become increasingly critical. Batching LLM requests is a critical step in scheduling the inference jobs...
484. Experimentally Extending Quantum Kernel Learning to Quantum Data by NMR ​
Author: Vivek Sabarad, Vishal Varma, T. S. Mahesh
Published: 8/19/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, physics.app-ph
arXiv:2412.09557v3 Announce Type: replace-cross Abstract: Quantum kernel learning (QKL) promises efficient machine learning by encoding feature maps onto exponentially large Hilbert spaces inherent in quantum systems. Using the liquid-state nuclear magnetic resonance (NMR) platform, we implement and...
485. Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities ​
Author: Qirun Dai, Dylan Zhang, Jiaqi W. Ma, Hao Peng
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2501.12147v2 Announce Type: replace-cross Abstract: Selecting appropriate training data is crucial for instruction fine-tuning of large language models (LLMs), which aims to (1) elicit strong capabilities, and (2) achieve balanced performance across different tasks. Influence-based methods sho...
486. Thinking Outside the (Gray) Box: A Context-Based Score for Assessing Value and Originality in Neural Text Generation ​
Author: Giorgio Franceschelli, Mirco Musolesi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG
arXiv:2502.13207v4 Announce Type: replace-cross Abstract: Despite the increasing use of large language models for creative tasks, their outputs often lack diversity. Common solutions, such as sampling at higher temperatures, can compromise the quality of the results. Dealing with this trade-off is s...
487. Bringing Generative Learning to Representation Learning: Self-Supervised Transfer Learning as Distribution Matching ​
Author: Yuling Jiao, Wensen Ma, Defeng Sun, Hansheng Wang, Yang Wang
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, stat.ME
arXiv:2502.14424v3 Announce Type: replace-cross Abstract: Most self-supervised learning objectives defend against collapse but leave the target representation law unspecified. We formulate representation learning as Distribution Matching (DM), learning an augmentation-invariant encoder whose induced...
488. Helios 2.0: A Robust, Ultra-Low Power Gesture Recognition System Optimised for Event-Sensor based Wearables ​
Author: Prarthana Bhattacharyya, Joshua Mitton, Ryan Page, Owen Morgan, Oliver Powell, Benjamin Menzies, Gabriel Homewood, Kemi Jacobs, Paolo Baesso, Taru Muhonen, Richard Vigars, Louis Berridge
Published: 8/19/2026, 4:00:00 AM
Categories: cs.HC, cs.CV, cs.LG
arXiv:2503.07825v3 Announce Type: replace-cross Abstract: We present an advance in wearable technology: a mobile-optimized, real-time, ultra-low-power event camera system that enables natural hand gesture control for smart glasses, dramatically improving user experience. While hand gesture recogniti...
489. Diffusion-model approach to flavor models: A case study for $S_4^\prime$ modular flavor model ​
Author: Satsuki Nishimura, Hajime Otsuka, Haruki Uchiyama
Published: 8/19/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-th
arXiv:2504.00944v2 Announce Type: replace-cross Abstract: We propose a numerical method of searching for parameters with experimental constraints in generic flavor models by utilizing diffusion models, which are classified as a type of generative artificial intelligence (generative AI). As a specifi...
490. On Stopping Times of Power-one Sequential Tests: Tight Lower and Upper Bounds ​
Author: Shubhada Agrawal, Ashwin Ram, Aaditya Ramdas
Published: 8/19/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.ML, stat.TH
arXiv:2504.19952v2 Announce Type: replace-cross Abstract: We present two general lower bounds for stopping times of sequential tests between arbitrary composite nulls $\mathcal P$ and alternatives $\mathcal Q$. The first lower bound is for the ``Wald setting'' where the type-1 error level $\alpha$ a...
491. Multilook Coherent Imaging: Theoretical Guarantees and Algorithms ​
Author: Xi Chen, Soham Jana, Christopher A. Metzler, Arian Maleki, Shirin Jalali
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, eess.IV
arXiv:2505.23594v2 Announce Type: replace-cross Abstract: Multilook coherent imaging is a widely used technique in applications such as digital holography, ultrasound imaging, and synthetic aperture radar. A central challenge in these systems is the presence of multiplicative noise, commonly known a...
492. FuseLIP: Multimodal Embeddings via Early Fusion of Discrete Tokens ​
Author: Christian Schlarmann, Francesco Croce, Nicolas Flammarion, Matthias Hein
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2506.03096v2 Announce Type: replace-cross Abstract: Contrastive language-image pre-training aligns features of text-image pairs in a common latent space via distinct encoders for each modality. While this approach achieves impressive performance in several zero-shot tasks, it cannot natively h...
493. Lightweight Deep Learning-Based Channel Estimation for RIS-Aided Extremely Large-Scale MIMO Systems on Resource-Limited Edge Devices ​
Author: Muhammad Kamran Saeed, Ashfaq Khokhar, Shakil Ahmed
Published: 8/19/2026, 4:00:00 AM
Categories: cs.IT, cs.CV, cs.LG, cs.NI, math.IT
arXiv:2507.09627v3 Announce Type: replace-cross Abstract: Next-generation wireless technologies such as 6G aim to meet demanding requirements such as ultra-high data rates, low latency, and enhanced connectivity. Extremely Large-Scale MIMO (XL-MIMO) and Reconfigurable Intelligent Surface (RIS) are k...
494. Projection-based multifidelity linear regression for data-scarce applications ​
Author: Vignesh Sella, Julie Pham, Karen Willcox, Anirban Chaudhuri
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.CE, cs.LG
arXiv:2508.08517v2 Announce Type: replace-cross Abstract: Surrogate modeling for systems with high-dimensional quantities of interest remains challenging, particularly when training data are costly to acquire. This work develops multifidelity methods for multiple-input multiple-output linear regress...
495. Encoding of musical structures in hidden units of restricted Boltzmann machines ​
Author: Mutsumi Kobayashi, Hiroshi Watanabe
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS
arXiv:2509.04899v4 Announce Type: replace-cross Abstract: Restricted Boltzmann machines (RBMs) are energy-based models originating from statistical physics, in which hidden units mediate the probability distribution of high-dimensional visible configurations. In this study, we use symbolic music as ...
496. SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA ​
Author: Haozhou Xu, Dongxia Wu, Matteo Chinazzi, Ruijia Niu, Rose Yu, Yi-An Ma
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2509.25459v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) show promise in generating long-form scientific explanations that synthesize evidence and connect multiple factors. However, in long-form scientific question answering, LLMs often hallucinate, producing unsupporte...
497. Comprehensive language-image pre-training for 3D medical image understanding ​
Author: Tassilo Wald, Ibrahim Ethem Hamamci, Yuan Gao, Sam Bond-Taylor, Harshita Sharma, Maximilian Ilse, Cynthia Lo, Olesya Melnichenko, Anton Schwaighofer, Noel C. F. Codella, Maria Teodora Wetscherek, Klaus H. Maier-Hein, Panagiotis Korfiatis, Valentina Salvatelli, Javier Alvarez-Valle, Fernando P'erez-Garc'ia
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2510.15042v3 Announce Type: replace-cross Abstract: In the 3D medical image domain, vision-language pre-training is used to create vision-language encoders (VLEs) that can support radiologists by retrieving patients with similar abnormalities, predicting likelihoods of abnormality, or, with do...
498. Iso-Riemannian Optimization on Learned Data Manifolds ​
Author: Willem Diepeveen, Melanie Weber
Published: 8/19/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DG
arXiv:2510.21033v3 Announce Type: replace-cross Abstract: We develop a theory of iso-Riemannian optimization for problems constrained to learned data manifolds, a setting in which classical Riemannian optimization - and Riemannian gradient descent in particular - can be poorly suited. That is, favor...
499. From Lab to Reality: A Practical Evaluation of Deep Learning Models and LLMs for Vulnerability Detection ​
Author: Chaomeng Lu, Bert Lagaisse
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.SE
arXiv:2512.10485v3 Announce Type: replace-cross Abstract: Vulnerability detection methods based on deep learning (DL) have shown strong performance on benchmark datasets, yet their real-world effectiveness remains underexplored. Recent work suggests that graph neural network-based and transformer-ba...
500. On Fibonacci Ensembles: An Alternative Approach to Ensemble Learning Inspired by the Timeless Architecture of the Golden Ratio ​
Author: Ernest Fokou'e
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2512.22284v2 Announce Type: replace-cross Abstract: Nature rarely reveals her secrets bluntly, yet in the Fibonacci sequence she grants us a glimpse of her quiet architecture of growth, harmony, and recursive stability \citep{Koshy2001Fibonacci, Livio2002GoldenRatio}. From spiral galaxies to t...
501. Gradient descent reliably finds depth- and gate-optimal circuits for generic unitaries ​
Author: Janani Gomathi, Alex Meiburg
Published: 8/19/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2601.03123v2 Announce Type: replace-cross Abstract: When the gate set has continuous parameters, synthesizing a unitary operator as a quantum circuit is, in principle, always possible using exact methods. However, efficiently finding depth- and gate-minimal circuits remains a major challenge. ...
502. Misconception Diagnosis From Student-Tutor Dialogue: Generate, Retrieve, Rerank ​
Author: Joshua Mitton, Prarthana Bhattacharyya, Digory Smith, Thomas Christie, Ralph Abboud, Simon Woodhead
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2602.02414v2 Announce Type: replace-cross Abstract: Timely and accurate identification of student misconceptions is key to improving learning outcomes and pre-empting the compounding of student errors. However, this task is highly dependent on the effort and intuition of the teacher. In this w...
503. MiDAS: A Multimodal Data Acquisition System and Dataset for Robot-Assisted Minimally Invasive Surgery ​
Author: Keshara Weerasinghe (MD), Seyed Hamid Reza Roodabeh (MD), Andrew Hawkins (MD), Zhaomeng Zhang, Zachary Schrader, Homa Alemzadeh
Published: 8/19/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG
arXiv:2602.12407v3 Announce Type: replace-cross Abstract: Background: Robot-assisted minimally invasive surgery (RMIS) research increasingly relies on multimodal data, yet access to proprietary robot telemetry remains a major barrier. We introduce MiDAS, an open-source, platform-agnostic system enab...
504. ReLoop: Structured Modeling and Behavioral Verification for Reliable LLM-Based Optimization ​
Author: Junbo Jacob Lian, Yujun Sun, Huiling Chen, Chaoyu Zhang, Hanzhang Qin, Chung-Piaw Teo
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG, math.OC
arXiv:2602.15983v3 Announce Type: replace-cross Abstract: Large language models (LLMs) can translate natural language into optimization code, but silent failures pose a critical risk: code that executes and returns solver-feasible solutions may encode semantically incorrect formulations---a feasibil...
505. mmWave Radar Aware Dual-Conditioned GAN for Speech Reconstruction of Signals With Low SNR ​
Author: Jash Karani, Adithya Chittem, Deepan Roy, Sandeep Joshi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SD, cs.LG
arXiv:2602.22431v2 Announce Type: replace-cross Abstract: Millimeter-wave (mmWave) radar captures are band-limited and noisy, making for difficult reconstruction of intelligible full-bandwidth speech. In this work, we propose a two-stage speech reconstruction pipeline for mmWave using a Radar-Aware ...
506. Random Quadratic Form on a Sphere: Synchronization by Common Noise ​
Author: Maximilian Engel, Anna Shalova
Published: 8/19/2026, 4:00:00 AM
Categories: math.PR, cs.LG, math.DS
arXiv:2603.06187v2 Announce Type: replace-cross Abstract: We introduce the Random Quadratic Form (RQF): a stochastic differential equation which formally corresponds to the gradient flow of a random quadratic functional on a sphere. While the one-point dynamics of the system is a Brownian motion and...
507. Mixture of experts architectures for machine learning interatomic potentials ​
Author: Yuzhi Liu, Duo Zhang, Anyang Peng, Weinan E, Linfeng Zhang, Han Wang
Published: 8/19/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, physics.comp-ph
arXiv:2603.07977v3 Announce Type: replace-cross Abstract: Machine Learning Interatomic Potentials (MLIPs) enable accurate large-scale atomistic simulations, yet improving their expressive capacity efficiently remains challenging. Here we systematically investigate Mixture-of-Experts (MoE) and Mixtur...
508. HiAP: A Multi-Granular Stochastic Auto-Pruning Framework for Vision Transformers ​
Author: Andy Li, Aiden Durrant, Milan Markovic, Georgios Leontidis
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2603.12222v3 Announce Type: replace-cross Abstract: Vision Transformers require significant computational resources and memory bandwidth, severely limiting their deployment on resource-constraint hardware. Most structured pruning methods reduce theoretical cost effectively, yet they typically ...
509. Less Data, Faster Convergence: Goal-Driven Data Optimization for Multimodal Instruction Tuning ​
Author: Rujie Wu, Haozhe Zhao, Hai Ci, Yizhou Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2603.12478v3 Announce Type: replace-cross Abstract: Multimodal instruction tuning is often compute-inefficient because training budgets are spread across large mixed image-video pools whose utility is highly uneven. We present Goal-Driven Data Optimization (GDO), a framework that computes six ...
510. Deep Learning for BioImaging: What Are We Really Learning? ​
Author: Ivan Svatko, Maxime Sanchez, Ihab Bendidi, Gilles Cottrell, Auguste Genovesio
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2603.13377v2 Announce Type: replace-cross Abstract: Representation learning has driven major advances in natural image analysis by enabling models to acquire high-level semantic features. In microscopy imaging, however, it remains unclear what current representation learning methods really lea...
511. Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought ​
Author: Zichen Xie, Wenxi Wang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2603.18334v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) increasingly assist secure software development, their ability to meet the rigorous demands of Rust program verification remains unclear. Existing evaluations treat Rust verification as a black box, assessing m...
512. SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs ​
Author: Yadi Cao, Sicheng Lai, Jiahe Huang, Yang Zhang, Zach Lawrence, Rohan Bhakta, Izzy F. Thomas, Mingyun Cao, Chung-Hao Tsai, Zihao Zhou, Yidong Zhao, Hao Liu, Alessandro Marinoni, Alexey Arefiev, Rose Yu
Published: 8/19/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.AI, cs.DC, cs.LG
arXiv:2603.20253v4 Announce Type: replace-cross Abstract: Evaluating LLM agents for scientific tasks has focused on token costs while ignoring tool-use costs like simulation time and experimental resources. As a result, metrics like pass@k become impractical under realistic budget constraints. To ad...
513. Camera-Agnostic Pruning of 3D Gaussian Splats via Descriptor-Based Beta Evidence ​
Author: Peter Fasogbon, Ugurcan Budak, Patrice Rondao Alface, Hamed Rezazadegan Tavakoli
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2603.21933v2 Announce Type: replace-cross Abstract: The pruning of 3D Gaussian splats is essential for reducing their complexity to enable efficient storage, transmission, and downstream processing. However, most of the existing pruning strategies depend on camera parameters, rendered images, ...
514. FactReview: Evidence-Grounded Peer Review with Execution-Based Claim Verification ​
Author: Ling Yue, Chaoqian Ouyang, Hang Xu, Ruijun Huang, Yuchen Liu, Libin Zheng, Wei Liu, Shaowu Pan, Shimin Di, Min-Ling Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2604.04074v4 Announce Type: replace-cross Abstract: Large language model (LLM)-based reviewing systems typically assess manuscripts in isolation, leaving literature- and code-dependent claims difficult to verify. We present FactReview, an audit pipeline that extracts review-relevant claims, gr...
515. Deterministic Adam-Inspired Methods with Accelerated Convergence Rate ​
Author: Yaxin Yu, Long Chen, Zeyi Xu
Published: 8/19/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2604.08742v2 Announce Type: replace-cross Abstract: Adam is widely used, but its convergence theory remains incomplete even in the deterministic full-batch setting because momentum and adaptive preconditioning are tightly coupled. For smooth convex objectives, we split the momentum variable th...
516. Morphology-Conditioned World Model for Cross-Embodiment Quadrupedal Locomotion ​
Author: Mohamad H. Danesh, Chenhao Li, Amin Abyaneh, Anas Houssaini, Kirsty Ellis, Glen Berseth, Marco Hutter, Hsiu-Chin Lin
Published: 8/19/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2604.08780v2 Announce Type: replace-cross Abstract: World models promise a paradigm shift in robotics, where an agent learns the physics of its environment once and then acquires behaviors efficiently. Yet the learned dynamics models at their core are typically morphology locked. In legged loc...
517. Transferable FB-GNN-MBE Framework for Potential Energy Surfaces: Data-Adaptive Transfer Learning in Deep Learned Many-Body Expansion Theory ​
Author: Siqi Chen, Zhiqiang Wang, Yili Shen, Xianqi Deng, Xi Cheng, Cheng-Wei Ju, Jun Yi, Guo Ling, Dieaa Alhmoud, Hui Guan, Zhou Lin
Published: 8/19/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG
arXiv:2604.09320v5 Announce Type: replace-cross Abstract: Mechanistic understanding and rational design of complex chemical systems depend on fast and accurate predictions of electronic structures beyond individual building blocks. However, if the system exceeds hundreds of atoms, first-principles q...
518. Differentially Private Verification of Distribution Properties ​
Author: Elbert Du, Cynthia Dwork, Pranay Tankala, Linjun Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.DS, cs.CC, cs.LG
arXiv:2604.10819v2 Announce Type: replace-cross Abstract: A recent line of work initiated by Chiesa and Gur and further developed by Herman and Rothblum investigates the sample and communication complexity of verifying properties of distributions with the assistance of a powerful, knowledgeable, but...
519. MoRFI: Monotonic Sparse Autoencoder Feature Identification ​
Author: Dimitris Dimakopoulos, Shay B. Cohen, Ioannis Konstas
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2604.26866v2 Announce Type: replace-cross Abstract: Large language models (LLMs) acquire most of their factual knowledge during the pre-training stage, through next token prediction. Subsequent stages of post-training often introduce new facts outwith the parametric knowledge, giving rise to h...
520. RETO: A Rotary-Enhanced Transformer Operator for High-Fidelity Prediction of Automotive Aerodynamics ​
Author: Bojun Zhang, Huiyu Yang, Yunpeng Wang, Yuntian Chen, Yuanwei Bin, Rikui Zhang, Jianchun Wang
Published: 8/19/2026, 4:00:00 AM
Categories: eess.IV, cs.LG
arXiv:2605.00062v3 Announce Type: replace-cross Abstract: Rapid aerodynamic evaluation is crucial for modern vehicle design, yet existing neural operators struggle to capture intricate spatial correlations. We propose the rotary-enhanced transformer operator (RETO), a novel neural solver featuring a...
521. Leakage-Audited Benchmarking Reveals Limited Evidence for Cross-Subject Auditory-Evoked EEG Vowel Perception Decoding ​
Author: Xiaoyang Li, Zeyan Tao
Published: 8/19/2026, 4:00:00 AM
Categories: eess.SP, cs.CL, cs.CV, cs.LG, cs.SD, q-bio.NC
arXiv:2605.00865v3 Announce Type: replace-cross Abstract: We tested whether auditory-evoked EEG supports subject-independent five-vowel perception decoding when trial identity, model identity, prediction provenance, and participant-level inference are controlled within a single benchmark. We reconst...
522. Evolving Ensemble of Agents ​
Author: Zongmin Yu, Liu Yang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG
arXiv:2605.09018v4 Announce Type: replace-cross Abstract: We introduce the Evolving Ensemble of Agents (EvE), a decentralized framework that organizes existing, highly capable coding agents into a live, co-evolving system for algorithmic discovery. Rather than reinventing the wheel within the ``LLMs...
523. The Observable Wasserstein Distance ​
Author: Edivaldo Lopes dos Santos, Leandro Vicente Mauri, Washington Mio, Tom Needham
Published: 8/19/2026, 4:00:00 AM
Categories: math.MG, cs.LG
arXiv:2605.09916v2 Announce Type: replace-cross Abstract: We introduce the observable Wasserstein distance, a framework for deriving lower bounds on the Wasserstein distance between probability measures on Polish metric spaces, designed to bypass the computational intractability of exact optimal tra...
524. AgentMV: A State-Guided Multi-Agent Framework for Budget-Aware Music Video Generation ​
Author: Huimin Wang, Chang Xia, Leilei Ouyang, Yongqi Kang, Yu Fu, Yuqi Ouyang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MA
arXiv:2605.10723v2 Announce Type: replace-cross Abstract: Generating a complete music video from a song requires more than synthesizing visually plausible clips for individual lyric prompts. A practical system must maintain long-range visual consistency, coordinate recurring motifs, synchronize edit...
525. An Agentic AI Framework with Large Language Models and Chain-of-Thought for UAV-Assisted Logistics Scheduling with Mobile Edge Computing ​
Author: Hanwen Zhang, Dusit Niyato, Wei Zhang, Xin Lou, Malcolm Yoke Hean Low
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2605.13221v2 Announce Type: replace-cross Abstract: In cloud manufacturing, unmanned aerial vehicles (UAVs) can support both product collection and mobile edge computing (MEC). This joint operation forms a hybrid scheduling problem, where physical logistics decisions are coupled with computati...
526. Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting ​
Author: Amritansh Maurya, Navjot Singh, Mohammed Javed, Omar Moured
Published: 8/19/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CV, cs.LG
arXiv:2605.20254v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown promising results on NLP tasks, however, their performance on tabular data still needs research attention, because Table Question-Answering (TQA) requires precise cell retrieval and multi-step structure...
527. The Economics of Model Collapse: Equilibrium, Welfare, and Optimal Provenance Subsidies in Synthetic Data Markets ​
Author: Gustav Olaf Yunus Laitinen-Fredriksson Lundstr"om-Imanov
Published: 8/19/2026, 4:00:00 AM
Categories: econ.GN, cs.CY, cs.LG, q-fin.EC
arXiv:2605.20279v2 Announce Type: replace-cross Abstract: Generative artificial intelligence is rapidly transforming the supply side of training data: an increasing share of new tokens, images, and structured records is produced by previous-generation models rather than by human originators. Recursi...
528. The Economics of AI Inference: Inflation Dynamics, Welfare Costs, and Optimal Monetary Policy under the Inference-Cost Phillips Curve ​
Author: Gustav Olaf Yunus Laitinen-Fredriksson Lundstr"om-Imanov
Published: 8/19/2026, 4:00:00 AM
Categories: econ.GN, cs.LG, q-fin.EC
arXiv:2605.20281v2 Announce Type: replace-cross Abstract: We develop a unified microeconomic and monetary theory of artificial intelligence inference costs and their pass-through to inflation, welfare, and optimal monetary policy. We introduce the Inference-Cost Phillips Curve (ICPC), an augmented N...
529. Practical and Optimal Algorithm for Linear Contextual Bandits with Rare Parameter Updates ​
Author: Sanghoon Yu, Min-hwan Oh
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2606.00984v2 Announce Type: replace-cross Abstract: We study linear contextual bandits under rare parameter updates: the learner may incorporate reward feedback into its parameter estimate only at a small number of update times, while still observing contexts online and selecting actions seque...
530. A Quantitative Approximation Framework for Flow Distillation in Diffusion Models ​
Author: Weiguo Gao, Ming Li, Lei Shi, Hanfei Zhou
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2606.03820v2 Announce Type: replace-cross Abstract: We develop a quantitative framework for diffusion distillation by viewing few step sampling as approximation through compositions of learned flow maps. For trajectory distillation of the probability flow ODE, we show that low noise multimodal...
531. Reward hacking in physical reinforcement learning revealed by turbulent drag reduction ​
Author: Giorgio Maria Cavallazzi, Miguel P'erez-Cuadrado, Alfredo Pinelli
Published: 8/19/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG
arXiv:2606.06227v3 Announce Type: replace-cross Abstract: Reinforcement-learning controllers optimise specified rewards, but in physical systems those rewards often capture only part of the true control objective. Three mechanisms through which this mismatch can produce apparent success without phys...
532. Computationally tractable robust differentially private mean estimation ​
Author: Kelly Ramsay
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML
arXiv:2606.12654v2 Announce Type: replace-cross Abstract: We develop a new, differentially private mean estimator called the balloon mean. The main features of the balloon mean are that it is computationally tractable and enjoys robustness to outlying observations. It is based on an iterative clippi...
533. Interpreting "Interpretability" and Explaining "Explainability" in Machine Learning in Physics ​
Author: Rikab Gambhir, Luisa Lucie-Smith, Jesse Thaler
Published: 8/19/2026, 4:00:00 AM
Categories: physics.data-an, astro-ph.GA, cs.LG, hep-ph
arXiv:2606.26228v2 Announce Type: replace-cross Abstract: We review the concepts of interpretability and explainability as they apply to machine learning in physics. We define interpretability as concerning the structural transparency of a model (the ability to understand or approximate its inner wo...
534. DanceOPD: On-Policy Generative Field Distillation ​
Author: Wei Zhou, Xiongwei Zhu, Zelin Xu, Bo Dong, Lixue Gong, Yongyuan Liang, Meng Chu, Leigang Qu, Lingdong Kong, Wei Liu, Tat-Seng Chua
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG
arXiv:2606.27377v3 Announce Type: replace-cross Abstract: Modern image generation demands a single model that unifies diverse capabilities, including text-to-image (T2I), local editing, and global editing. However, these capabilities are rarely naturally aligned and often conflict. For instance, edi...
535. Sampling the Schwinger Model with Gauge-Equivariant Diffusion ​
Author: Octavio Vega, Aida X. El-Khadra
Published: 8/19/2026, 4:00:00 AM
Categories: hep-lat, cond-mat.str-el, cs.LG
arXiv:2606.27481v2 Announce Type: replace-cross Abstract: We present a first study of a diffusion-based approach to accelerated sampling of the $N_f = 2$ lattice Schwinger model. Our work is inspired by recent and growing successes in developing such generative models for ensemble generation in LFT ...
536. On the Role of Directionality in Structural Generalization ​
Author: Zichao Wei
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.02307v2 Announce Type: replace-cross Abstract: Several SLOG test categories explicitly involve directional distinctions (modifier position shifts, argument extraction positions), yet AM-Parser, the previous SOTA, uses an AM algebra whose operations do not encode direction. We redesign the...
537. Statistical Adversaries: Natural Backdoor-like Adversarial Features in Clean Vision Datasets ​
Author: Paul K. Mandal, Pavan Reddy, Tristan Malatynski
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CR, cs.LG
arXiv:2607.05516v2 Announce Type: replace-cross Abstract: Model-specific adversarial attacks have been extensively studied. We study a different failure mode: naturally occurring statistical signals in vision data that can behave as backdoor-like triggers without being maliciously inserted. We call ...
538. Lesioned Multimodal Language Models Reproduce Aphasic Picture-Naming Patterns ​
Author: Yong Yang, Xiang Guan, Sophie Arheix-Parras, Saeed Ahmadi, Roger Newman-Norlund, Leonardo Bonilha, Christopher Rorden, Julius Fridriksson, Rutvik H. Desai, Srihari Nelakuditi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.11621v2 Announce Type: replace-cross Abstract: Aphasia following stroke commonly produces systematic naming errors with characteristic profiles, but whether general-purpose language models not designed for clinical simulation can reproduce these patterns remains untested. We investigated ...
539. The Many Senses of Visual Similarity: A Text-Prompted Image Perceptual Metric ​
Author: Sheng-Yu Wang, Yotam Nitzan, Aaron Hertzmann, Jun-Yan Zhu, Eli Shechtman, Alexei A. Efros, Richard Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.18237v2 Announce Type: replace-cross Abstract: Human visual similarity judgments are context-dependent. For example, two images may be similar in shape but distinct in color. Existing perceptual similarity metrics, however, collapse these nuances into a single scalar value, offering no me...
540. Priors learned from legacy reconstructions inherit undetectable overconfidence ​
Author: Ali Siahkoohi, Sina Alemohammad
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.21721v3 Announce Type: replace-cross Abstract: Where truths are scarce (e.g., seismic and medical imaging), a prior for an ill-posed inverse problem is trained on an archive of legacy reconstructions---an older method's outputs---and its uncertainty is treated as data-driven. In the popul...
541. Fast Trainable Multilinear Bases for Image Compression ​
Author: Shiwen An, Zhongyi Ni, Huanhai Zhou, Jin-Guo Liu
Published: 8/19/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, math.OC, quant-ph
arXiv:2608.00053v2 Announce Type: replace-cross Abstract: The Discrete Fourier Transform, the Discrete Cosine Transform, and their block-wise variants underpin most deployed image and video codecs. Their effectiveness rests on three properties: they run in near-linear time (linear up to a polylogari...
542. UOT-IR: Structured Routing of High-Polyphony Symbolic Music into Fixed-Budget Representations ​
Author: Ziyue Kang, Nan Nan, Chenhao Lin, Xiaohong Guan
Published: 8/19/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG
arXiv:2608.00576v2 Announce Type: replace-cross Abstract: High-polyphony symbolic music is increasingly used in generation, analysis, and arrangement, yet many downstream tasks require bounded representations with fixed tracks or slots. Converting richly orchestrated scores into compact forms is the...
543. How Benchmarks and Evaluation Protocols Shape Conclusions in Provenance-Based Intrusion Detection ​
Author: Lorenzo Guerra, Thomas Chapuis, Guillaume Duc, Pavlo Mozharovskyi, Van-Tam Nguyen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.01454v2 Announce Type: replace-cross Abstract: Provenance-based intrusion detection systems (PIDS) frequently report strong performance, but the conclusions drawn from these results can be highly sensitive to benchmarking choices and evaluation protocols. We investigate this dependency by...
544. Non-KKT Accumulation in Entropic Mirror Descent ​
Author: Kuangyu Ding, Kim-Chuan Toh
Published: 8/19/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DS
arXiv:2608.01658v3 Announce Type: replace-cross Abstract: For mirror descent generated by a Legendre kernel, perhaps one of the most basic question in optimization is this: must every accumulation point of a bounded mirror descent sequence be Karush--Kuhn--Tucker (KKT) stationary under proper stepsi...
545. Integrating spectral and morphological plant features with decision-tree models for early-season cotton biomass and nitrogen status estimation from multi-year UAV data ​
Author: Vaishali Swaminathan, Nithya Rajan, J Alex Thomasson, Amrit Shrestha, Karem Meza Capcha, Robert Hardin, Pramod Pokhrel
Published: 8/19/2026, 4:00:00 AM
Categories: eess.IV, cs.LG
arXiv:2608.07801v2 Announce Type: replace-cross Abstract: Precision nitrogen (N) management (PNM) for cotton requires in-season monitoring of crop growth parameters and N status indicators to decide fertilizer timing, placement, and application rates for optimal canopy development and yield. This st...
546. The Spectral Neuron ​
Author: Alex Shtoff
Published: 8/19/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.08003v2 Announce Type: replace-cross Abstract: As machine learned models increase in complexity and expressive power, features of simpler models, such as intrinsic coefficient transparency and control over the shape of the modeled function are lost. On the one edge of the spectrum we have...
547. Population-Scalable Multi-Agent World Modeling ​
Author: Renjie Zhao, Yuxiang Wu, Mingyu Zhang, Jiaxin Li, Sisi Li, Yimin Sheng, Tianxi Tan, Zhenkai Zhang, Jianyi Zhu, Yong-Lu Li
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.08600v2 Announce Type: replace-cross Abstract: World models have recently achieved impressive progress in visual prediction and interactive generation, but extending them to multi-agent environments introduces a fundamental scalability challenge. Existing methods generally assume a fixed ...
548. When Do Anchor-Based Pointwise LLM Rerankers Help? Retriever Quality, Statistical Scope, and Anchor Design ​
Author: Utshab Kumar Ghosh, Shubham Chatterjee
Published: 8/19/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.10528v3 Announce Type: replace-cross Abstract: Anchor-based pointwise LLM reranking scores each candidate against a shared reference passage to recover cross-document context at pointwise cost. We study when this actually helps, using GCCP/PAGC as a representative method. Our study is rep...
549. When Self-Consistency Backfires: Majority Vote Hurts the Majority of Hard Science Problems for Small LLMs ​
Author: Utkarsh Bahuguna
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.11403v2 Announce Type: replace-cross Abstract: Self-consistency via majority vote reduces per-problem accuracy on most GPQA Diamond problems for small instruction-tuned models: 56.6% of problems for Qwen2.5-7B and 65.7% for Llama-3-8B. The obvious remedy is a verifier-free confidence gate...
550. Cross-Corpus Evaluation of Generalizable Vulnerability Detection in IoT Firmware ​
Author: Sadib Hassan Rumman, Md. Shariful Islam, Md Rayhanur Rahman
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.11492v2 Announce Type: replace-cross Abstract: IoT firmware vulnerability detection is constrained by ecosystem heterogeneity, resource-limited platforms, and benchmark quality limitations. Existing datasets are often synthetic or general-purpose and lack human-verified, contamination-scr...
551. One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL ​
Author: Simon Yu, Nicholas Tomlin, Marwa Abdulhai, Ximing Lu, Derek Chong, Abe Hou, Dilara Soylu, Sergey Levine, Christopher D. Manning, Weiyan Shi
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.12253v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that this approach systematically fails to generalize, and trace the failure to simulator collaps...
552. Excess Separability: Nuisance-Controlled Residual-Stream Probing for Benchmark Contamination Detection ​
Author: Florian Braun
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.12652v2 Announce Type: replace-cross Abstract: Benchmark contamination is diagnosed with n-gram overlap, likelihood-based membership inference, or canary strings, and each needs something usually unavailable: the training corpus, a well-chosen test statistic, or foresight at release. A re...
553. MobileMem: Learning from a Year of Mobile Experiences ​
Author: Xinle Deng, Yida Xue, Xiangyuan Ru, Yijun Chen, Buqiang Xu, Mingjun Mao, Xinjie Liu, Haoming Xu, Shuofei Qiao, Mengru Wang, Chen Jiang, Yuchen Eleanor Jiang, Lizhong Wang, Jason Wang, Li Zeng, Haofen Wang, Guilin Qi, Huajun Chen, Ningyu Zhang
Published: 8/19/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA, cs.MM
arXiv:2608.13606v2 Announce Type: replace-cross Abstract: The next generation of AI agents is increasingly moving beyond systems that answer isolated questions toward persistent personal assistants that can understand, remember, and continuously learn from users' experiences. Such assistants require...
554. GALA: Generation-Aware Cross-Modal Alignment for Text-to-Time-Series Synthesis ​
Author: Haochen Zhang, Gengwei Zhang, Laura Yao, Nicholas Konz, Tianlong Chen
Published: 8/19/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.13741v2 Announce Type: replace-cross Abstract: Synthesizing time series from natural language is emerging as the most expressive form of controllable time series generation. However, existing text-conditioned generators either take caption embeddings frozen from off-the-shelf text encoder...
555. Pairton: Iterative Reconstruction of Short-Lived Particles ​
Author: Andreas Hermansen, Chris Scheulen, Tobias Golling
Published: 8/19/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-ex
arXiv:2608.14278v2 Announce Type: replace-cross Abstract: We present Pairton, an iterative framework for reconstructing short-lived particles in high-energy collision events. By formulating particle reconstruction as a masked prediction process over graph structures, Pairton learns conditional distr...