Skip to content

arXiv cs.LG - 2026-08-18 ​

555 items collected.


1. Learning Discrete Riemannian Metrics for Physical Fields with Cochain-Frame Equivarianc ​

Author: Dongzhe Zheng, Christine Allen-Blanchette
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14556v1 Announce Type: new Abstract: Physical fields on meshes require a separation between topology and geometry: conservation laws are topological and should be exact, while geometry, material response, and anisotropic coupling must be learned from data. Existing neural surrogates often...

📖 Read original article


2. Forward Pass Domain Adaptation (Without Cross-Layer Backpropagation) ​

Author: Rivaan Patil, Simon Dennis, Hao Guo, Kevin Shabahang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14563v1 Announce Type: new Abstract: Forward-Pass-Only MLP training (FPO) adapts large language models without a backward pass through the model body, achieving 2.7--3.2x the throughput of standard fine-tuning at ~40% less peak training memory, while leaving off-domain benchmarks within s...

📖 Read original article


3. Coarse-to-Fine Multi-Resolution Diffusion Models for Trajectory Generation in Urban Systems ​

Author: Wen Ye, Muyan Weng, Chuizheng Meng, Hao Niu, Yizhou Zhang, Yan Liu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14570v1 Announce Type: new Abstract: Understanding human mobility is critical for a wide range of urban applications, including traffic management, epidemic control, and urban planning. However, due to privacy concerns, the availability of large-scale public trajectory data remains limite...

📖 Read original article


4. Geometry Is Not Robustness: A Trajectory-Level Study of PGD Evaluation ​

Author: Dhairysheel Durgule
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14594v1 Announce Type: new Abstract: Projected Gradient Descent (PGD) is widely used to evaluate adversarial robustness, typically via final adversarial accuracy, which does not capture model behaviour throughout the attack. Recent work proposes trajectory-level diagnostics, such as loss ...

📖 Read original article


5. DumpsterCluster: From Dumpster Diving to Serving LLaMA-70B on $60 GPUs ​

Author: Zeyu Cao, Xuan Guo, Cheng Zhang, Cheuk Hang Lau, Ilia Shumailov, Yiren Zhao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR

arXiv:2608.14614v1 Announce Type: new Abstract: As AI datacenters retire functional GPUs, vast quantities of still capable accelerators enter secondary markets. This paper investigates whether these retired GPUs can find a productive afterlife to form a DumpsterCluster that can serve modern LLM infe...

📖 Read original article


6. Calibrated Trust, Not Sharper Prediction: An Empirical Test of Uncertainty Fusion ​

Author: Surya Saka
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.14617v1 Announce Type: new Abstract: A recurring proposal in legal AI is to improve case-outcome prediction by fusing uncertainty tools (evidence graphs with belief propagation, sequential Bayesian odds updating, Dempster-Shafer combination, and conformal prediction) into one pipeline. We...

📖 Read original article


7. PIKFNO: An Interpretable Neural Operator Based on Physics Informed Kernel Function ​

Author: Yuan Guo, Hanshu Chen, Zhuojia Fu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14619v1 Announce Type: new Abstract: This work proposes a new interpretable neural operator framework, termed the Physics Informed Kernel Function Neural Operator (PIKFNO), which explicitly incorporates physics informed kernel functions derived from governing equations into the neural ope...

📖 Read original article


8. Explaining Reinforcement Learning Decisions in Self-adaptive Systems ​

Author: Jasmina Gajcin, Juan C. Rosero, Ivana Dusparic
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14620v1 Announce Type: new Abstract: Reinforcement Learning (RL) has been extensively used in autonomous and self-* systems, but RL policies, especially deep RL ones relying on neural networks, lack transparency and are difficult to understand. This can lead to diminished user trust, and ...

📖 Read original article


9. Metaplasticity as adaptive gradient preconditioning for incremental learning ​

Author: Isabelle Aguilar, Zayn Andre Zainal, Omid Kavehei
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14634v1 Announce Type: new Abstract: Biological intelligence naturally prevents catastrophic forgetting through Complementary Learning Systems (CLS) theory, a macroscopic consolidation process driven at the local level by synaptic metaplasticity: the continuous, history-dependent neuromod...

📖 Read original article


10. Fractional Optimizers Meet Fractal Activation Functions: An Empirical Study of Multi-Scale Optimization in Neural Network ​

Author: Sebastian Raubitzek, Georg Goldenits, Sebastian Schrittwieser, Philip K"onig, Kevin Mallinger
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14636v1 Announce Type: new Abstract: Fractional optimization methods and fractal activation functions are two independent directions for improving neural network training. Fractional optimizers extend first-order optimization through fractional derivatives and memory effects, whereas frac...

📖 Read original article


11. Early Cycle Charge Trajectory Generative Prediction and Full Life Cycle Health Management of Iron-Chromium Flow Batteries Based on FlowBD-E1 ​

Author: Suyang Zhuang, Zekun Jiang, Tianhang Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ME

arXiv:2608.14637v1 Announce Type: new Abstract: Long-duration stationary energy storage requires batteries whose degradation can be detected before substantial capacity loss has accumulated. Iron-chromium redox flow batteries are attractive for this role because they use abundant and low-cost active...

📖 Read original article


12. Randomly initialized autoencoders: fixed points and edge-of-chaos ​

Author: Leonid Berlyand, Roman Sarapin, Yitzchak Shmalo, Victor Slavin, Sasha Sodin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math.PR, math.ST, stat.TH

arXiv:2608.14638v1 Announce Type: new Abstract: In this paper we study autoencoders, a special class of deep neural nets (DNNs) whose performance can be characterized via their fixed points. This perspective naturally raises questions of existence, stability, and basins of attraction of these fixed ...

📖 Read original article


13. Valid Per-Field Selective Risk Control for Document Extraction: Three Failure Modes, a Validity Ladder, and When Conditioning Pays ​

Author: Bhaskar Gurram
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.14639v1 Announce Type: new Abstract: Per-field accept/review with selective risk at most alpha -- accept a field only if the error rate among accepted fields is controlled -- is the trust contract document-extraction systems need, and the natural procedure silently violates it on real doc...

📖 Read original article


14. BDIP-Net: Dual-Interaction Graph Learning for Property Prediction of Bilayer Materials ​

Author: An Vuong, Chen Zhao, Jin Hu, Shui-Qing Yu, Xintao Wu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.AI

arXiv:2608.14640v1 Announce Type: new Abstract: Stacked bilayer materials exhibit rich stacking-dependent properties driven by the interplay between strong intra-layer bonding and weak inter-layer van der Waals interactions. The computational discovery of such materials is challenging because accura...

📖 Read original article


15. Training and Evaluating Ethical Reinforcement Learning Agents on Per-Episode Distributions ​

Author: Prabhjyot Singh, Majid Ghasemi, Mark Crowley
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14642v1 Announce Type: new Abstract: Reinforcement Learning (RL) agents trained on a single reward signal exploit the gap between the designed reward and the intended behavior. This is particularly a problem when we are trying to imbue ethical behavior into RL agents. An agent can look et...

📖 Read original article


16. DUET: Dual-Teacher On-Policy Distillation via Same-Weight Disagreement for Prohibition Compliance ​

Author: Zihan Li, Feifei Li, Wenhui Que
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.14644v1 Announce Type: new Abstract: Real-world LLM deployments increasingly rely on runtime-injected prohibitions--enterprise policies, PII redlines, tool boundaries--that vary per request and per tenant. Conventional post-training is structurally ill-suited: SFT hides the violation sign...

📖 Read original article


17. Efficient Neural-Network-Based High-Resolution Radiative Transfer for CO___ Retrieval, and Application to Interferometric Sensing ​

Author: Jordan Lontsi Tedongmo (CB), Yann Ferrec (CB, IFUMI), Laurence Croiz'e (CB, IFUMI), Pablo Mus'e (CB, IFUMI), Gabriele Facciolo (CB), Andr'es Almansa (MAP5 - UMR 8145, IFUMI)
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2608.14645v1 Announce Type: new Abstract: Studying climate change requires reducing uncertainties in CO2 and CH4 emission estimates to better distinguish anthropogenic from natural sources, which motivates spaceborne measurements with improved revisit frequency and spatial coverage. In this co...

📖 Read original article


18. iFuzz-Meta: An Interpretable Fuzzy Learning Framework Bridging Top-Down and Bottom-Up Knowledge Integration ​

Author: Xiaowei Jiang, Daniel Leong, Beining Cao, Nan Zhou, Yingtao Ren, Yu-Cheng Chang, Thomas Do, Chin-Teng Lin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC

arXiv:2608.14646v1 Announce Type: new Abstract: Interpretable representation learning remains a key challenge in modern neural computation, particularly when models are expected not only to perform but also to explain their reasoning. This paper introduces iFuzz-Meta, an interpretable fuzzy rule-bas...

📖 Read original article


19. SMOPD: Selective Token-Entropy Masking for Dirty-History Multi-Turn On-Policy Self-Distillation ​

Author: Chenyang Jiang, Changhan Huang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14647v1 Announce Type: new Abstract: Dirty-history rollouts make multi-turn on-policy self-distillation (OPSD) brittle: once a student emits an erroneous intermediate reply, later turns are conditioned on that reply, and uniform distillation can spend loss on tokens that carry little corr...

📖 Read original article


20. Discrete Diffusion Language Models Are Training-Free Multi-Label Classifiers ​

Author: Pawan Kumar
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14649v1 Announce Type: new Abstract: We present dLLM-SetScore, a training-free method that uses discrete masked-diffusion language models for multi-label text classification. For each candidate label, it asks a short yes/no question and compares the probabilities of the two answer tokens ...

📖 Read original article


21. Paired Exact-Reset Evaluation of a Prediction-Derived Medium-to-Full World-Model Cascade ​

Author: Malo de Pastor
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.14650v1 Announce Type: new Abstract: Existing adaptive-inference and world-action-model systems use cheap-stage outputs or predicted futures to allocate additional computation. We study a narrower question: under paired exact-reset physical outcomes, can a Medium-derived interface predict...

📖 Read original article


22. Pushing the Limits of High-Resolution Weather Forecasting through Data Scaling ​

Author: Yang Zhao, Peisong Niu, Tian Zhou, Ziqing Ma, Guanlong Ma, Rong Jin, Huiling Yuan, Liang Sun
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.14652v1 Announce Type: new Abstract: The development of 0.1$^{\circ}$ global weather forecasting models based on machine learning (ML) is constrained by the limited availability of high-resolution data, as decades of reanalysis are only available at 0.25$^{\circ}$ resolution. While existi...

📖 Read original article


23. Do Uncertainty Signals Help? A Systematic Study of Uncertainty-Aware Decoding with Rollback Mechanisms ​

Author: Xianzong Wu, Xiaohong Li, Yuejun Guo, Xinyang Liu, Tianlin Li, Junjie Wang, Qiang Hu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14653v1 Announce Type: new Abstract: Prediction uncertainty is a widely adopted metric for quantifying model confidence, with downstream applications spanning model explanation, data selection, and prediction rollback. Despite its demonstrated utility, the potential of uncertainty quantif...

📖 Read original article


24. FedImp: Enhancing Federated Learning Convergence with Impurity-Based Weighting ​

Author: Hai Anh Tran, Cuong Ta, Truong X. Tran
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14654v1 Announce Type: new Abstract: Federated Learning (FL) is a collaborative paradigm that enables multiple devices to train a global model while preserving local data privacy. A major challenge in FL is the non-Independent and Identically Distributed (non-IID) nature of data across de...

📖 Read original article


25. Diagnosing and Mitigating Perception-Decision Misalignment in Omni-LLMs via Modality Subspace Activation ​

Author: Hongbo Jiang, Jie Li, Yunhang Shen, Tianyu Xie, Pingyang Dai
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.14655v1 Announce Type: new Abstract: Omni-Large Language Models (Omni-LLMs) power complex multi-modal reasoning in applications like World Action Models and autonomous agents. However, their strong performance often masks a profound Perceptual-Decision Misalignment (PDM), where decisions ...

📖 Read original article


26. P2E-VQ: ECG-linked representation augmentation for PPG via discrete patch retrieval ​

Author: Zhongli Wu, Zhuangzhi Gao, He Zhao, Feixiang Zhou, Fu Wang, Jinru Ding, Yuankai Wang, Hongyi Qin, Gregory Y. H. Lip, Bil Kirmani, Yalin Zheng
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14656v1 Announce Type: new Abstract: Photoplethysmography (PPG) is widely used in consumer wearables because of its low cost and ease of acquisition. However, unlike electrocardiography (ECG), PPG measures peripheral pulse dynamics rather than cardiac electrical activity, limiting its abi...

📖 Read original article


27. LUNG-KGMM: Knowledge-Guided Multimodal Learning for Lung Cancer Incidence Prediction ​

Author: Chunlei Yang, Shuyan Li, Zhong Cao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.14657v1 Announce Type: new Abstract: Early identification of lung cancer risk is critical for timely intervention, yet existing prediction models are limited by their reliance on single data modalities and their inability to leverage structured clinical knowledge. We propose LUNG-KGMM, a ...

📖 Read original article


28. pico-type: A 1.5M-Parameter Byte-Level Multi-Head Content Classifier ​

Author: Gautam Kishore
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.IR

arXiv:2608.14658v1 Announce Type: new Abstract: We introduce pico-type, a byte-level multi-head content classifier with approximately 1.5 million parameters that simultaneously predicts seven content properties from raw UTF-8 bytes in a single forward pass. Operating directly at the byte level -- no...

📖 Read original article


29. Ring-based Spatial Transformer: Learning Non-linear Spatial Interactions between Building Distribution and Pedestrian Flow ​

Author: Shun Nakayama, Takahiro Kanamori, Wanglin Yan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY

arXiv:2608.14660v1 Announce Type: new Abstract: This study proposes a ring-based SpatialTransformer to learn how building uses at different distances from a railway station interact to generate pedestrian flow. Concentric ring buffers at 100-meter intervals up to 800 meters were defined around 100 r...

📖 Read original article


30. An automatic-differentiation framework for time-lapse electrical resistivity tomography inversion of hydrologic dynamics ​

Author: Pu Yang, Zhengyang Fang, Yuxin Liu, Xuan Su, Deshan Feng, Hang Chen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14661v1 Announce Type: new Abstract: Time-lapse electrical resistivity tomography (TL-ERT) provides spatially distributed information on subsurface hydrologic changes. However, inversion of long monitoring sequences is computationally demanding. Modifying the data misfit, regularization, ...

📖 Read original article


31. In-Context Learning to Assess Built Environment Impacts on Perceived Neighborhood Walkability Among Mobility-impaired Older Adults ​

Author: Houhao Liang, Kresimir Friganovic, Joanne Kua, Noor Hafizah Ismail, Su Su, Bryan Yijia Tan, Navrag B. Singh, Panos Mavros
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.AP

arXiv:2608.14663v1 Announce Type: new Abstract: As global populations age, enhancing neighborhood walkability through inclusive urban design is important for mitigating built environment (BE) barriers that discourage physical activity and social participation among older adults. This study investiga...

📖 Read original article


32. Quantifying Depth Sufficiency in Residual Neural Networks: A First-Order Criterion ​

Author: Zeyu Liu, Jinhao Zhang, Yunquan Zhang, Guangming Tan, Xiang Gao, Fangming Liu, Daning Cheng
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14664v1 Announce Type: new Abstract: How can we determine whether a trained neural network is already deep enough? We study this under a fixed function-preserving residual-growth protocol specifying insertion locations, residual families, zero-output initializations, and zero-state first-...

📖 Read original article


33. When Does the Best Sampling Temperature Rise with the Budget? Sufficient Conditions for Pass@k ​

Author: Changsu Jeong (Independent Researcher)
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14665v1 Announce Type: new Abstract: The temperature that maximizes pass@$k$ is often low for a small sampling budget and higher for a large budget. This pattern has been reported from Codex through recent multi-sample inference studies. It is not an algebraic property of pass@$k$: as Slo...

📖 Read original article


34. ARGUS: Attention-Guided Transformers for Scalable Person Identification Using Wi-Fi Telemetry ​

Author: Nayan Sanjay Bhatia, Pranay Kocheta, Yuhan Li, Katia Obraczka
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.14670v1 Announce Type: new Abstract: Passive, device-free person identification offers an alternative to camera- and wearable-based biometrics, yet existing wireless approaches rely largely on gait or activity cues and are rarely evaluated at scale. In this paper, we present \emph{Argus},...

📖 Read original article


35. Take it Personally: The Limits of General SSL Representations for Real-Life PPG Emotion Detection ​

Author: Dominika Kunc, Przemys{\l}aw Kazienko, Stanis{\l}aw Saganowski
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14675v1 Announce Type: new Abstract: While Self-Supervised Learning (SSL) effectively extracts general representations from noisy, unconstrained physiological signals such as photoplethysmography (PPG), its suitability for highly subjective tasks remains unproven. In this work, we evaluat...

📖 Read original article


36. RouteTS: Frequency-Time Routing for Time Series Forecasting ​

Author: Gaofeng Lin, Lei Duan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.14682v1 Announce Type: new Abstract: Real-world time series inherently intertwine global periodic structures with localized non-stationary variations. Existing approaches process these heterogeneous dynamics within a single computational domain, incurring fundamental limitations: time-dom...

📖 Read original article


37. One Score, Two Decisions: Selective Prediction on the Rare-Disease Tail ​

Author: Zhaoyang Jiang, Zhizhong Fu, Yunsoo Kim, Zicheng Li, Xuanqi Peng, Fei Teng, Jiacong Mi, Honghan Wu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14683v1 Announce Type: new Abstract: Given a patient's clinical findings, a diagnostic system ranks possible diseases and must decide when to endorse its first prediction or defer it for review. This decision is usually made by thresholding the top score. Selective prediction over ranked ...

📖 Read original article


38. Mitigating Rubric Interference in LLM Judges via On-Policy Self-Distillation ​

Author: Dingyao Yu, Tong Zhang, Yutao Mou, Yunxiao Zhang, Wei Ye, Shikun Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14684v1 Announce Type: new Abstract: LLM judges increasingly evaluate responses against fine-grained rubric checklists. When a sample requires multiple rubrics, current methods typically assess each in a separate inference call. Evaluating all rubrics in a single pass is a natural alterna...

📖 Read original article


39. Rethinking Reverse KL as Adaptive Entropy Distillation ​

Author: Shizhen Li, Zhiyu Shen, Yuyin Lu, Yunhe Pang, Jielin Song, Yanghui Rao, Fu Lee Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.14685v1 Announce Type: new Abstract: Knowledge distillation (KD) is widely used to transfer the capabilities of large language models (LLMs) to smaller students, but existing objectives often struggle to balance faithful imitation and robust generation. In particular, existing methods mai...

📖 Read original article


40. A Reproducibility Study of Partial Residual Ablations in Pre-LN Transformers ​

Author: Pratikkumar Babariya
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14689v1 Announce Type: new Abstract: Residual connections are a fundamental component of transformer architectures, yet the roles of the attention and feed-forward residual pathways remain poorly understood when considered independently. This paper presents a reproducibility study of part...

📖 Read original article


41. The Quantum Shortcut: Complex Phase-State Dynamics Reduce the Optimization Steps of Sequence Models ​

Author: Ahmed Nebli, Hadi Saadatdoorabi, Christopher Keibel, Kevin Yam
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, quant-ph

arXiv:2608.14691v1 Announce Type: new Abstract: Sequence models are conventionally distinguished by their backbone, the mechanism that routes information across positions, such as attention or recurrence. This paper varies a choice that is prior to the backbone and shared by nearly all current model...

📖 Read original article


42. Tail-Aware Top-$k$ On-Policy Distillation ​

Author: Huipeng Huang, Hongxin Wei
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14728v1 Announce Type: new Abstract: On-policy distillation (OPD) has emerged as an effective paradigm for transferring knowledge between language models, where a student is trained to align its next-token distribution with the teacher's along its own trajectories. To provide dense superv...

📖 Read original article


43. A Novel Fourier Feature Network for Solving Partial Differential Equations ​

Author: Qihong Yang, Zhijie Su, Yangtao Deng, Qiaolin He
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14733v1 Announce Type: new Abstract: Building on the foundation of single-hidden-layer neural networks, Fourier Feature Networks (FENs) are proposed, which incorporate Fourier features using $\cos$, $\sin$, or a combination of both. Similar to Extreme Learning Machines (ELMs), FENs employ...

📖 Read original article


44. Unraveling the Size Determination Mechanism of Nanocrystal Synthesis via Interpretable Neural Networks ​

Author: Kai Gu, Haizheng Zhong
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.AI

arXiv:2608.14734v1 Announce Type: new Abstract: Deep learning models of nanocrystal synthesis enable the prediction of size and shape by encoding precursors and reaction conditions. However, their black-box nature hinders gaining deep insights into the underlying synthetic mechanisms. Here, we devel...

📖 Read original article


45. Generative Learning of Separatrices ​

Author: Ellis R. Crabtree, Dimitris G. Giovanis, Anastasia Georgiou, George Datseris, Ioannis G. Kevrekidis
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math.DS, stat.ML

arXiv:2608.14743v1 Announce Type: new Abstract: The identification and reconstruction of the boundaries separating basins of attraction in multistable, multidimensional dynamical systems presents a fundamental challenge in computational dynamics. These structures govern transition pathways and other...

📖 Read original article


46. Iterative Refinement Diffusion for Super-Resolved Data Assimilation of Multiscale Physical Systems ​

Author: Mrigank Dhingra, Ramchandran Muthukumar, Rebecca Willett, Omer San
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn

arXiv:2608.14744v1 Announce Type: new Abstract: Recovering high-resolution states from sparse, low-resolution observations is a central challenge in scientific machine learning and data assimilation. Classical data assimilation exploits temporal information through forecast-analysis cycles, but ofte...

📖 Read original article


47. WANDR: A Benchmark for Wide and Deep Research ​

Author: Vitaliy Polshkov, Marcin Pitera, Jeremy Yang, Kirill Priemko, Maksim Gaiduk, Aleksandr Nikolenko, Denis Bykov, Clare Southern, Denis Yarats, Jerry Ma
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14747v1 Announce Type: new Abstract: WANDR (Wide ANd Deep Research) is a benchmark of 500 realistic, challenging data-collection tasks for research agents. Each task requires a system to discover a large set of entities that satisfy specified criteria (breadth), investigate each entity th...

📖 Read original article


48. Real-Time State-of-Health Estimation and Online Degradation Prognosis from Partial Battery Discharge Using Physics-Informed Neural Networks ​

Author: Bego~na Ispizua, Serio Gil-L'opez, Leire Arrizabalaga, Ibai La~na
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14764v1 Announce Type: new Abstract: With the increasing integration of renewable energy sources, energy storage systems have become essential, making the accurate estimation of their State of Health (SOH) and degradation behavior critical. In this work, we propose a physics-informed deep...

📖 Read original article


49. ER-KANs: Efficient and Robust Kolmogorov-Arnold Networks for Data-Scarce Scientific Machine Learning ​

Author: Harshil Lodhiya
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.14773v1 Announce Type: new Abstract: The efficient-KAN literature---covering Chebyshev, wavelet, and radial-basis-function variants of the original Kolmogorov-Arnold Network---has been benchmarked almost entirely on clean data. We show that this choice conceals a large capability differen...

📖 Read original article


50. p-Spin Glass Network Efficient Single-Batch Continual Learning ​

Author: Vladimer Khasia
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14774v1 Announce Type: new Abstract: Modern sequence models heavily rely on massive memory footprints and large-batch stochastic optimization, barriers that restrict sample efficiency and continual learning. We introduce the $p$-Spin Glass Network, a novel architecture that overcomes thes...

📖 Read original article


51. NRCD: An Open Database of Collegiate Running with Unified Performance Standardization ​

Author: Jonathan A. Karr Jr., Ryan M. Fryer, Ben Darden, Nicholas Pell, Kayla Ambrose, Evan Hall, Ramzi K. Bualuan, Nitesh V. Chawla
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CY, cs.IR

arXiv:2608.14776v1 Announce Type: new Abstract: Collegiate running in the United States generates thousands of race results annually in cross country and track and field, yet no large-scale dataset has been publicly available for research. Existing websites such as Athletic.net, MileSplit, and TFRRS...

📖 Read original article


52. Is Grokking a Loss of Normal Hyperbolicity of the Interpolation Manifold? ​

Author: Suvinava Basak
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14803v1 Announce Type: new Abstract: A recent line of work recasts the post-memorization phase of grokking as constrained optimization: once a network interpolates the training set, weight decay drives a slow drift along the zero-loss manifold toward lower norm. In the language of dynamic...

📖 Read original article


53. Disentangling Homophily and Rarity: Explaining Failure in Graph Neural Networks ​

Author: Preben M. Ness, Fariz Ikhwantri, Dusica Marijan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14823v1 Announce Type: new Abstract: Are heterophilic nodes in a graph harder to classify because they are heterophilic or because they are rare? Some existing work frames classification of such nodes as a subgroup generalisation problem, where a model performs well on the majority group ...

📖 Read original article


54. M-LINKX: Multiview Graph Learning for Brain Cognitive Disease Detection ​

Author: An Phan, Yufei Jin, Xingquan Zhu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14847v1 Announce Type: new Abstract: Electroencephalogram (EEG) is a non-invasive and relatively low-cost procedure that measures brain electricity for the detection of cognitive diseases. EEG-based classification of dementia-related conditions, including Alzheimer's disease (AD), mild co...

📖 Read original article


55. Can Neural Networks Learn by Experimenting on Themselves? Self-Interventional Learning from Functional Consequences to Predictive Self-Knowledge ​

Author: Micha{\l} Tomaszewski
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14894v1 Announce Type: new Abstract: Machine-learning systems usually model external data, while their internal functional organization is analyzed by external observers. This work introduces Self-Interventional Learning (SIL), in which a neural system perturbs its own functional structur...

📖 Read original article


56. Degeneracy Counting Quantum Algorithm using Decoherence ​

Author: Malay Marut Das, Mark A. Novotny, Yaroslav Koshka
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, quant-ph

arXiv:2608.14941v1 Announce Type: new Abstract: Counting the global optima of a classical optimization problem is a #P-hard task. We develop the canonical thermal pure quantum (CTPQ) state-based degeneracy counting (CTPQsd#) algorithm that determines the number of global optima of a classical optimi...

📖 Read original article


57. PathFinder: Joint Decompositions of Linked Multimodal Datasets ​

Author: Ying-Qiu Zheng, Alex Fung, Stephen M Smith, Rogier B Mars, Saad Jbabdi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, eess.IV, q-bio.QM, stat.ML

arXiv:2608.14951v1 Announce Type: new Abstract: Low-rank matrix decompositions can uncover patterns and structure in data and have a number of different applications across many disciplines. Extensions to "joint" low-rank decompositions have been proposed to link datasets from different modalities. ...

📖 Read original article


58. LLM-based Framework for Generating and Verifying Parallel DEVS Statecharts ​

Author: Vamsi Krishna Vasa, Hessam S. Sarjoughian, Edward J. Yellig
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.LO

arXiv:2608.14956v1 Announce Type: new Abstract: The development of models demands sound modeling and simulation knowledge as well as domain knowledge. Every model should accurately represent a system's dynamics and be verifiable. Toward this objective, this research introduces an agentic PDEVS-LLM f...

📖 Read original article


59. Command-Space Counterfactual Explanations for Pareto-Conditioned Reinforcement Learning ​

Author: Joanikij Chulev, Hendrik Baier
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC

arXiv:2608.14963v1 Announce Type: new Abstract: Pareto Conditioned Networks learn multiple multi-objective reinforcement learning behaviours by conditioning a single policy on a desired return command. However, the local mapping from command and state to action remains opaque. We propose command-spa...

📖 Read original article


60. A Physiology-Informed Digital Twin Framework for Simulating Liver Health Progression ​

Author: Sumaiya Afroz Mila, Sandip Ray
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.14969v1 Announce Type: new Abstract: We present a physiology-informed digital twin of the human liver designed for longitudinal simulation of liver function and early-stage disease progression. The model, referred to as HEPATWIN, integrates key hepatic processes, including carbohydrate, l...

📖 Read original article


61. Do Geometry-Aware Positional Encodings Help Transformers in Spatial Imperfect-Information Games? ​

Author: Wenji Fu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.14982v1 Announce Type: new Abstract: Transformers applied to spatial imperfect-information games must represent map geometry while tracking hidden entities through time. We ask whether geometry-aware positional encodings improve these capabilities, without claiming a new positional encodi...

📖 Read original article


62. Lipschitz Bandits with Arbitrary Feedback Delays ​

Author: Yuhao Liu, Yu Chen, Longbo Huang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15036v1 Announce Type: new Abstract: The Lipschitz bandit problem extends the traditional multi-armed bandit framework to continuous action spaces by assuming that the reward functions satisfy a Lipschitz condition. This work investigates Lipschitz bandits under arbitrary feedback delays,...

📖 Read original article


63. Certifying Compressed Language Models: An Audit and a Statistical Toolkit ​

Author: Amogh Singh
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15046v1 Announce Type: new Abstract: A fraction of a point of benchmark accuracy is the usual evidence that a compressed model is equivalent to its original. That quantity is least informative when two models are most alike: a net delta is what survives cancellation between opposing per-i...

📖 Read original article


64. Online Convex Optimization with Dueling Feedback ​

Author: Yiyang Lu, Hareshkumar Jadav, Mohammad Pedramfar, Ranveer Singh, Vaneet Aggarwal
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15050v1 Announce Type: new Abstract: We study online convex optimization with dueling (pairwise comparison) feedback, where the learner observes only a binary preference between two queried points. While dueling feedback is well understood in discrete or stochastic settings, the adversari...

📖 Read original article


65. A Unified Mamba--MoE Surrogate for Closed-Loop Simulation and Measurement-Window Forecasting of Inverter Transients ​

Author: Haoguang Wang, Huy Hoang Le, Akhila Kandivalasa, Christian Moya, Marcos Netto, Guang Lin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2608.15051v1 Announce Type: new Abstract: This paper proposes a Mamba surrogate model with mixture-of-experts (MoE) routing to represent the transient dynamics of inverter-based resources. A Mamba surrogate model is a predictive machine learning model built on the Mamba architecture. MoE routi...

📖 Read original article


66. GATTA: Graph Active Learning with Test-Time Augmentation ​

Author: Zsombor B'anfi, Andr'as G'ezsi, Andr'as Formanek
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15084v1 Announce Type: new Abstract: Test-time augmentation (TTA) has proven effective for improving model robustness and uncertainty estimation in computer vision, yet its application to graph-structured data remains largely unexplored. We introduce GATTA (Graph Active Learning with Test...

📖 Read original article


67. EMASAM: a Computationally Efficient Sharpness-Aware Minimization via EMA-Guided Perturbations ​

Author: Tanapat Ratchatorn, Masayuki Tanaka
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.15105v1 Announce Type: new Abstract: Recent progress in optimization research has highlighted the sharpness of the loss landscape as a key factor in narrowing the generalization gap. Motivated by this insight, Sharpness-Aware Minimization (SAM) was proposed as a training strategy that enh...

📖 Read original article


68. Global Federated Learning Strategies for Building Efficient Personalized Models ​

Author: Seongyoon Kim
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15107v1 Announce Type: new Abstract: Federated learning (FL) is a practical framework that can train models on distributed user data while guaranteeing data privacy; however, due to heterogeneity in which each user has a different data distribution, problems frequently arise where both gl...

📖 Read original article


69. Probability-Preserving Transformer for the Time-Dependent Schr\"odinger Equation ​

Author: Mushtaq Ali, Muzamil Tariq, Niaz Ali Khan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, quant-ph

arXiv:2608.15112v1 Announce Type: new Abstract: Solving the time-dependent Schr"odinger equation (TDSE) via traditional numerical methods is computationally intensive. Transformer models offer a compelling alternative, but standard implementations rely on soft constraints that cannot rigorously gua...

📖 Read original article


70. Decision-Driven Regularization: A Blended Model for Learning and Optimization ​

Author: Gar Goei Loke, Qinshen Tang, Yangge Xiao, Xun Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.15124v1 Announce Type: new Abstract: In contextual optimization, the decision-maker seeks optimal decisions to minimize a cost function, that varies based on observed features. This context is common in many business applications ranging from on-demand delivery and retail operations to po...

📖 Read original article


Author: Alexander L. Strehl
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15146v1 Announce Type: new Abstract: We revisit Tesauro's TD-Gammon for backgammon money games in the setting of no evaluation-time search. Both checker play and cube action (use of the doubling cube) are learned from scratch via self-play reinforcement learning (RL), with minimal hand-co...

📖 Read original article


72. FinFraudBench: A Heterogeneous Graph Benchmark for Financial Fraud Detection ​

Author: Yixuan Chen, Hongyu Zhan, Jie Sheng, Weiyu Han, Shuai Chen, Tianyi Zhang, Xiao Tan, Jun Xia
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15177v1 Announce Type: new Abstract: The increasing complexity of digital financial systems has reshaped financial fraud detection from isolated transaction classification into relational risk reasoning over interconnected financial entities. This shift has motivated graph-based fraud det...

📖 Read original article


73. MiNO: Cotangent-bundle propagator learning for PDEs ​

Author: Gnankan Landry Regis N'guessan, Bum Jun Kim
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.NA, math.NA

arXiv:2608.15187v1 Announce Type: new Abstract: Scientific machine learning for partial differential equations commonly targets solution fields, as in physics-informed neural networks, or solution maps, as in neural operators. We study a third target: the propagator itself, a phase and amplitude in ...

📖 Read original article


74. Structuring Semantic Embeddings for Principle Evaluation: A Prototype-Guided Contrastive Learning Approach ​

Author: Che Shen, Junwei Su, Lingpeng Kong, Chuan Wu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15224v1 Announce Type: new Abstract: Reliable post-hoc evaluation asks whether already generated text satisfies a target criterion after generation. In this paper we study a focused frozen-embedding setting using principle-evaluation proxy tasks: toxicity detection, fine-grained emotion c...

📖 Read original article


75. Learning reshapes power-law anisotropy in internal representations ​

Author: Asahi Nakamuta, Jun-nosuke Teramae
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.15239v1 Announce Type: new Abstract: Power-law anisotropy in internal representations has been observed across a wide range of biological and artificial neural systems, from state-of-the-art language models to the mouse cerebral cortex. This anisotropy is a key geometric property of high-...

📖 Read original article


76. Earth Observation Foundation Models for Terrestrial Ecohydrology: From Representation Learning to Process Inference ​

Author: Yi Yu, Jian Peng, Yucheng Lin, Trevor F. Keenan, Thomas F. A. Bishop
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, physics.bio-ph

arXiv:2608.15282v1 Announce Type: new Abstract: Earth observation foundation models (EOFMs) are emerging as reusable representation frameworks for data-driven retrieval, prediction and process modelling within ecohydrology, which integrate EO, meteorological forcing and process models to characteris...

📖 Read original article


77. No Task Fails Every Time: Why One-Shot Audits Are Structurally Blind to Agent Damage ​

Author: Shiven Khurdi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15286v1 Announce Type: new Abstract: We introduce AgentRelBench, an environment-agnostic reliability instrument that computes ground-truth, severity-priced damage from database state diffs across repeated runs, with no LLM in the measurement path, demonstrated on EnterpriseOps-Gym. Across...

📖 Read original article


78. MAPLE: MoE Adaptive Plug-and-play Layer-wise Expert allocation ​

Author: Lie Li, Wen Li, Junxiao Shen, Gusheng Hu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15299v1 Announce Type: new Abstract: Sparsely-activated Mixture-of-Experts (MoE) Transformers universally fix the same number of routed experts across all layers, a convention that ignores the well-documented heterogeneity in layer-wise redundancy. We demonstrate that this uniformity is s...

📖 Read original article


79. Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning ​

Author: Alexandre L. M. Levada
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, stat.ML

arXiv:2608.15313v1 Announce Type: new Abstract: In this paper, we propose SHOPCA (Shape Operator-based Principal Component Analysis), a novel method for unsupervised metric learning and dimensionality reduction that incorporates differential geometric information into the covariance structure of cla...

📖 Read original article


80. Spectral Rank Certification for Foundation Model Adapters ​

Author: Mohammed Ahnouch, Lotfi Elaachak
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ME

arXiv:2608.15351v1 Announce Type: new Abstract: Nominal LoRA rank is a design parameter; calibrated spectral evidence is a separate inferential quantity. This article develops a finite-sample framework for inferring effective rank structure in public foundation-model adapters. The theoretical core i...

📖 Read original article


81. SAPE: Sandwich Adapters for Parameter Efficiency in Large Language Model Fine-Tuning ​

Author: Mohammad Aref Jafari-Raddani, Morteza Mohajjel Kafshdooz
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15360v1 Announce Type: new Abstract: While Parameter-Efficient Fine-Tuning (PEFT) has substantially reduced the hardware cost of adapting Large Language Models (LLMs) by decreasing the number of trainable parameters, recent studies have sought to further improve PEFT through parameter sha...

📖 Read original article


82. Does 1/2-Tsallis-INF Also Work Well for Best-Arm Identification? ​

Author: Jingxin Zhan, Yuze Han, Zhihua Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15365v1 Announce Type: new Abstract: Regret minimization (RM) and best-arm identification (BAI) are two fundamental objectives in multi-armed bandits. Among regret-minimizing algorithms, $1/2$-Tsallis-INF is a canonical best-of-both-worlds FTRL algorithm: it achieves logarithmic pseudo-re...

📖 Read original article


83. Beyond Field Accuracy: Two-Axis Diagnosis of Inverse-PINN Parameter Error ​

Author: Yifan Zhang, Qian Tao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2608.15373v1 Announce Type: new Abstract: Inverse physics-informed neural networks (PINNs) can reconstruct a field accurately while returning an incorrect physical parameter. We introduce a two-axis post-training diagnosis that separates finite-sample resolution under a specified observation-a...

📖 Read original article


84. Every Expert Counts: ExactMoE for Memory-Efficient W4A16 Inference ​

Author: Amjad Saab
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15383v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) language models reduce arithmetic by activating only a small subset of experts per token, yet deployment still requires storing and moving the full expert bank. We present ExactMoE, an inference design that applies symme...

📖 Read original article


85. Look Before You Lift: Visual and Quantitative Diagnostics for Topological Deep Learning ​

Author: Mathilde Papillon, Guillermo Bern'ardez, 'Alvaro Ball'on Barreiro, Marco Montagna, R'emi Devaux, Antoine Jardin, Nina Miolane
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15388v1 Announce Type: new Abstract: Topological deep learning (TDL) methods rely on lifting raw data into higher-order discrete domains such as simplicial complexes, cell complexes, and hypergraphs. In practice, this lifting step is often treated as a black box: practitioners select a li...

📖 Read original article


86. Towards a theory of inference-time alignment with unknown rewards ​

Author: Steve Hanneke, Hongao Wang, Mingyue Xu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15402v1 Announce Type: new Abstract: Generative model alignment has received broad interest, and significant progress has been made in supervised fine-tuning and inference-time computation. Yet, alignment has remained poorly understood from a statistical learning perspective. We formulate...

📖 Read original article


87. FAST-DeepONet: Factor-Augmented Branch Representations for High-Dimensional PDE Inputs in the Small-Sample Regime ​

Author: Jiyong Kwon, Bongseok Kim, Guang Lin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15408v1 Announce Type: new Abstract: Deep operator networks can become statistically unstable when partial differential equation inputs are observed at thousands of strongly correlated sensors but only a small number of operator samples is available. We introduce FAST-DeepONet, a branch r...

📖 Read original article


88. Invariant Pretraining for Robust Code Representations ​

Author: Yifeng He, Yundi Xu, Christopher Castro Gaw Gonzalo, Zili Wang, Hao Chen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE

arXiv:2608.15412v1 Announce Type: new Abstract: Encoder-based code representation models remain widely deployed for discriminative tasks such as clone detection and code classification, where their small size and low inference cost are decisive. Their robustness, however, is fragile: under invariant...

📖 Read original article


89. SAGA: Structure-Attended Generative Action Embedding Model that encodes Multi-Surface User Action Sequences ​

Author: Tsz Fung Pang, Po Jen Chen, Nimish Ronghe, Farhad Farahani, Bo Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2608.15429v1 Announce Type: new Abstract: Prior embedding models for sequential recommendation typically operate within a homogeneous action space, limiting their ability to capture cross-surface behavioral signals spanning distinct behavioral domains. We present SAGA, a generative action embe...

📖 Read original article


90. Detecting Money Laundering in Rwandan Mobile Money: A Machine Learning Framework ​

Author: Emmanuel Nahimana, Ya'e Ulrich Gaba
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, q-fin.RM

arXiv:2608.15447v1 Announce Type: new Abstract: Mobile money has widened financial access across Sub-Saharan Africa and enlarged the surface for money-laundering and terrorism-financing (ML/TF) activity in ecosystems dominated by high-volume, low-value transactions. Rwanda is a case in point: severa...

📖 Read original article


91. Not All Attention Is Equal: A Quantitative Survey of the EEI Trade-off ​

Author: Aditya Singh
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15459v1 Announce Type: new Abstract: Attention mechanisms have driven machine learning for a decade, from neural machine translation to language models that do general-purpose reasoning. This survey covers four connected threads: their formulation for sequence-to-sequence tasks, adaptatio...

📖 Read original article


92. High-Dimensional Nonparametric Change-Point Detection via Low-Rank Degree-Three Density Projection ​

Author: Guoqing Zhang, Zhaixin Chen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15466v1 Announce Type: new Abstract: Distributional changes can be invisible to means and covariances yet appear in skewness, asymmetric interactions, or other third-order structure. We develop a nonparametric change-point method that retains every degree-at-most-three coefficient of a de...

📖 Read original article


93. Population Structure Analysis of an Inbred Population using Quantitative Shape Phenotyping from Stereo Retinal Photographs ​

Author: Li Tang, Michael D Abramoff
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, q-bio.QM

arXiv:2608.15471v1 Announce Type: new Abstract: The population structure of an inbred population of 781 people on Norfolk Island in the Pacific, 318 of which are descendants of the original Mutineers of the Bounty, is analyzed phenotypically using shape from stereo retinal fundus photographs. Three-...

📖 Read original article


94. Optimal Lower Bounds for Networked Information Aggregation ​

Author: Ambar Pal
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.15472v1 Announce Type: new Abstract: The problem of networked information aggregation, studied in Kearns et al. (2026), involves a group of learners situated on the vertices of a directed acyclic graph $G$, each learning a linear predictor $\widehat Y$ for a fixed random variable $Y$ give...

📖 Read original article


95. Measuring Structured Predictability in Neural Training Dynamics: A Cross-Regime Study ​

Author: Fanqi Wang, Weisheng Tang, Hairong Qi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15483v1 Announce Type: new Abstract: Modern deep networks are trained through long update trajectories, yet their temporal organization remains less systematically characterized than architectures, losses, or optimizers. We study short-horizon predictability as a measure of temporal redun...

📖 Read original article


96. QSMP: finding representative time series subsequences through Quick Shift+Matrix Profile ​

Author: Carlos H. Mendoza-Cardenas, Rogers F. Silva, Austin J. Brockmeier
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15492v1 Announce Type: new Abstract: Finding representative waveforms in long time series has scientific and practical value in many domains, as it enables summarization and visualization of large time series datasets, and downstream tasks like classification and forecasting. We present h...

📖 Read original article


97. PERO: Efficient Robust Post-Training Foundation Models for Encrypted Traffic Classification ​

Author: Wumei Du, Jiarong Wen, Kaiyu Zhang, Zi Yang, Yiqin Lv, Longfei Zhang, Dong Liang, Zheng Xie
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.15504v1 Announce Type: new Abstract: Encrypted traffic classification is vital for network security, yet real-world deployments are inherently sensitive to rare but high-loss errors such as misclassification of malicious traffic. The encrypted traffic foundation model, as a promising gene...

📖 Read original article


98. UniFed-VLM: Federated Instruction Tuning for Vision-Language Models with Multiple Heterogeneity ​

Author: Pengyu Wang, Baochen Xiong, Xiaoshan Yang, Yifan Xu, Zhang Qimeng, Haifeng Chen, Changsheng Xu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15516v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated strong performance in multimodal understanding and generation. However, fine-tuning of VLMs typically relies on centralized data, which raises privacy concerns in certain domains (e.g. healthcare). Federa...

📖 Read original article


99. Guaranteed Adaptive Modality Acquisition: When the Policy Chooses Its Own Calibration Group ​

Author: Melika Baghi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15520v1 Announce Type: new Abstract: A multimodal system may begin inference holding only some of its inputs and may acquire the rest at a cost. With adaptive acquisition, the policy determines which inputs are ultimately observed, so we state the guarantee conditional on that terminal in...

📖 Read original article


100. Spectral Saliency for Machine Unlearning ​

Author: Cedar Site Bai, Amber Yijia Zheng, Raymond A. Yeh, Brian Bullins
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15548v1 Announce Type: new Abstract: Machine unlearning (MU) aims to remove the influence of specific training data while preserving model utility. As the name suggests, MU can be viewed as the inverse of learning, using gradient-based updates to reduce the influence of a forget-set by co...

📖 Read original article


101. Amortised Post-Hoc Explanation with Exact Preservation for Dynamic Graph Anomaly Detectors ​

Author: Iyad Assaad Nekka, Hamida Seba, Walid Khaled Hidouci, Karima Amrouche
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15559v1 Announce Type: new Abstract: Anomaly detection in dynamic graphs underpins financial fraud analysis, intrusion detection, and platform integrity, where automated decisions require human-interpretable justifications. StrGNN, the strongest performer in recent benchmarks, produces no...

📖 Read original article


102. SchurQuant: Groupwise Discrete Optimization for Layer-Wise LLM Quantization ​

Author: Gunjun Lee, Sehwan Son, Younjoo Lee, Byungjun Kim, Jung Ho Ahn
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15567v1 Announce Type: new Abstract: Weight-only post-training quantization (PTQ) enables the deployment of large language models under tight memory budgets, but accuracy often collapses at 2-3 bits. Existing backpropagation-free PTQ optimizers have two limitations: group decisions ignore...

📖 Read original article


103. GraniKV: Asymmetric Granularity KV-Cache Paging for Multi-Agent Systems with Long Shared Prefix ​

Author: Jinhyun Jeon, Sungjoo Yoo
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15584v1 Announce Type: new Abstract: Production paged-serving engines apply uniform paging granularity to the KV cache, even though the two regions of a multi-agent workload have opposite storage requirements: a long shared prefix demands contiguity, while the per-request suffix demands f...

📖 Read original article


104. Quantum Models with Multi-Stage Training for Compositional Concept Generalization ​

Author: Mina Abbaszadeh, Matilda Karabina Moore, Mehrnoosh Sadrzadeh, Martha Lewis
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15601v1 Announce Type: new Abstract: Compositional Concept Generalization (CoCoGen), the ability to systematically recombine learned primitives in novel contexts, is a key challenge for multimodal learning. In this work, we provide a solution using a compositional model of meaning that se...

📖 Read original article


105. FluxBin: Flexible LUT-based Ultra-low-bit LLM Inference by Algorithm-Kernel Synergy ​

Author: Qingyao Yang, Runming Yang, He Xiao, Wendong Xu, Junyu Chen, Haobo Liu, Chenchen Ding, Ruihan Hu, Yik-Chung Wu, Ngai Wong
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15602v1 Announce Type: new Abstract: While binary quantization theoretically promises extreme compression and acceleration for Large Language Models (LLMs), existing research often overlooks the necessity of specialized hardware kernels, thus failing to unleash the full acceleration poten...

📖 Read original article


106. Benchmarking Quantum Machine Learning for Power-System Attack Detection: Evaluation Choices Decide the Outcome Before the Models Do ​

Author: Md Rezwanul Islam
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, stat.ML

arXiv:2608.15617v1 Announce Type: new Abstract: Machine-learning detectors for power-system cyberattacks are themselves attack surfaces, and quantum machine learning has been proposed for them. We benchmark fidelity-kernel SVMs and variational classifiers against six tuned classical models on public...

📖 Read original article


107. In Defense of OCTA: The Reconstruction-Utility Gap in OCT-to-OCTA Synthesis ​

Author: Michael Chertok, Alon Tiosano, Orly Gal-Or, Lior Kramarski, Einav Baharav Shlezinger, Irit Bahar, Lior Wolf
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15626v1 Announce Type: new Abstract: Optical coherence tomography angiography (OCTA) images retinal blood flow, giving capillary-perfusion and foveal-avascular-zone biomarkers that grade diabetic-retinopathy ischemia. Because OCTA hardware is less common than structural OCT, recent work s...

📖 Read original article


108. Sparse Prototype Code Underlies Classification and Prediction Across Modalities ​

Author: Yehonatan Avidan, Daniel D. Lee, Haim Sompolinsky
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, cs.AI, stat.ML

arXiv:2608.15632v1 Announce Type: new Abstract: Neural representations have become a central tool for studying the internal mechanisms of modern AI models, yet their complex high-dimensional structure makes them difficult to interpret. We show that classification tasks give rise to a universal repre...

📖 Read original article


109. Generalised Transportability via Causal Abstractions ​

Author: Yorgos Felekis, Paris Giampouras, Fabio Massimo Zennaro, Theodoros Damoulas
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15645v1 Announce Type: new Abstract: Transporting a causal conclusion from a source study population to a target one is a fundamental problem in causal inference. The theory of transportability provides a criterion for when this is possible: given experimental data from the source and obs...

📖 Read original article


110. Sequential Multimodal Evidence Optimization for Product Media Ranking in E-Commerce ​

Author: Prasenjit Dey, Frank McIntyre, Arnab Sinha
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15662v1 Announce Type: new Abstract: On modern e-commerce stores, customers consume ordered slates of heterogeneous product media, such as images, videos, and 3D renders, before making purchase decisions. Existing media-ranking systems often optimize myopic engagement proxies such as clic...

📖 Read original article


111. SubZero+: Efficient Zeroth-Order LLM Fine-Tuning via Large Learning Rates ​

Author: Ziming Yu, Shuyao Xiao, Xingyu Zhao, Sike Wang, Pan Zhou, Peiyu Zang, Xiangda Yan, Yongjie Yang, Jia Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15665v1 Announce Type: new Abstract: Zeroth-order (ZO) optimization enables backpropagation-free fine-tuning of large language models, but existing ZO methods suffer from high-variance gradient estimators, making convergence unstable and highly sensitive to learning rates. We propose SubZ...

📖 Read original article


Author: Zhongwei Yu, Yan Song, Xue Yan, Anjie Liu, Xingyu Lu, Yihang Chen, Huichi Zhou, Siyuan Guo, Luoyang Sun, Sihan Chen, Xiangning Yu, Jun Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15669v1 Announce Type: new Abstract: Scientific discovery often involves optimising expensive-to-evaluate objectives over vast, structured, and open-ended hypothesis spaces, such as molecules, protein sequences, and computer programs. Generative models such as large language models (LLMs)...

📖 Read original article


113. PL-Guard: Probabilistic Logic Reasoning for LLM Guardrails ​

Author: Satchit Chatterji, Shihan Wang, Giovanni Sileno, Erman Acar
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15673v1 Announce Type: new Abstract: Large language model guardrails can be viewed as policy-consistency problems: a system must determine which policy-relevant facts hold in a prompt-response pair and what those facts imply under a given policy. Common approaches, including policy prompt...

📖 Read original article


114. Learning Auditable Classifier Models: Source-Disjoint Tree Ensembles ​

Author: Srikumar Krishnamoorthy
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15725v1 Announce Type: new Abstract: Predictive models in clinical and regulated settings must be accurate and fully auditable. Tree ensembles deliver strong accuracy on tabular data, but their sequential boosting couples structure discovery with coefficient estimation, making compact per...

📖 Read original article


115. FirstDiff: One-Step Diffusion-Based Anomaly Detection for Multivariate Time Series via Initial Noise Prediction ​

Author: Ali Boudaghi, Alireza Nemati, Hadi Zare
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2608.15727v1 Announce Type: new Abstract: Diffusion models have recently shown strong potential for multivariate time-series anomaly detection by learning the distribution of normal data through iterative denoising. Existing diffusion-based approaches, however, typically perform anomaly detect...

📖 Read original article


116. TinyCast: Probabilistic Zero-Shot Forecasting with Computed Periodicity ​

Author: Armin Steinhauser
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15767v1 Announce Type: new Abstract: We introduce TinyCast, an attention-free zero-shot forecaster that emits a predictive distribution from 146,505 parameters, on the premise that at this size the periodic structure of a context is worth computing rather than learning. A zero-parameter s...

📖 Read original article


117. Temporal Graph Prototype-conditioned Conformal Prediction for Fraud Detection ​

Author: Xudong Chen, Shengbo Gong, Lu Cheng, Wei Jin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15768v1 Announce Type: new Abstract: Conformal prediction (CP) provides distribution-free coverage guarantees and has emerged as a principled tool for uncertainty quantification. In edge-level fraud detection on temporal interaction graphs, where false positives and false negatives both c...

📖 Read original article


118. Learning Stock Trading Policies via Barycenter-Based Adversarial Inverse Reinforcement Learning ​

Author: Arishi Orra, Himanshu Choudhary, Manoj Thakur
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.15770v1 Announce Type: new Abstract: Designing effective trading strategies using reinforcement learning remains challenging due to delayed and noisy rewards, poor exploration, and the difficulty of enforcing explicit risk constraints. In this work, we propose BRaG, a barycenter-based adv...

📖 Read original article


119. Routing Divergence Is Not Evidence of Behavioral Influence in Same-Weight MoE Self-Distillation ​

Author: Cedric Caruzzo, Donggeun Yoo, Tae Soo Kim
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.15787v1 Announce Type: new Abstract: Two Mixture-of-Experts (MoE) forward passes can share every weight yet route the same token through different experts. This creates a possible blind spot in same-weight self-distillation, where a demonstration-conditioned teacher supervises a query-onl...

📖 Read original article


120. CrevasseSeg: A Label-Efficient UAV Crevasse Segmentation Framework ​

Author: Steven Wallace, William D Harcourt, Richard Hann, Aiden Durrant, Somayajulu Sripada, Georgios Leontidis
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15790v1 Announce Type: new Abstract: Crevasse mapping from uncrewed aerial vehicle (UAV) imagery matters for glaciological research and for field safety in glaciated terrain. Yet, pixel-level annotation of glacier surfaces is costly and requires domain experts. We introduce CrevasseSeg, a...

📖 Read original article


121. Cross-Entropy Risk Estimation for Language Models: Inconsistency Must Be Dense, and the Holdout Method Is No Exception ​

Author: Hanti Lin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.15798v1 Announce Type: new Abstract: Language models are compared by their held-out per-token cross-entropy risk---the quantity scaling laws are fitted to. We show that it cannot be consistently estimated. Consistency, or convergence to the estimand, is defined relative to a \emph{possibl...

📖 Read original article


122. A Pre-Specified Construction-Confirmation Test of Operation-Level Causal Transfer Across Finite Isomorphic Symbolic Domains ​

Author: Xinyi Shan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15809v1 Announce Type: new Abstract: Behavioral accuracy, linear decodability, and successful activation interventions do not by themselves show that a model carries an operation-level structure from one symbolic domain to another. We ask a narrower question in finite isomorphic state spa...

📖 Read original article


123. KOALA: Koopman Operator Learning for WiFi-Based Anticipatory Hum ​

Author: Quang-Anh N. D., Duc Pham Minh, Thao Phuong Pham, Minh Anh Nguyen, Huan X. Nguyen, Tuan Dang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15815v1 Announce Type: new Abstract: WiFi Channel State Information (CSI) has emerged as a privacy-preserving alternative to cameras for human pose estimation. However, existing approaches treat pose inference as an instantaneous regression problem and do not model temporal dynamics, maki...

📖 Read original article


124. Second-Moment Memory in Coordinatewise Adam ​

Author: Jeonseong Kim
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15824v1 Announce Type: new Abstract: Adam retains a moving average of past squared gradients in its denominator, but the optimization cost of this memory is not well understood. We show that second-moment memory can itself suppress progress toward the optimum even under finite-variance st...

📖 Read original article


125. Self-Supervised Auxiliary Task Discovery for Stable Reinforcement Learning in Stock Trading ​

Author: Arishi Orra, Himanshu Choudhary, Manoj Thakur
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP, stat.ML

arXiv:2608.15841v1 Announce Type: new Abstract: Reinforcement learning has gained increasing attention as a data-driven approach for stock trading. However, learning a policy that is both profitable and stable remains challenging due to non-stationary market behaviour and noisy reward signals. Auxil...

📖 Read original article


126. Geometry of Forgetting: Representation Flux in Continual Learning ​

Author: Maksim A. Kazanskii
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.15854v1 Announce Type: new Abstract: Catastrophic forgetting remains a fundamental obstacle to continual learning, where neural networks lose previously acquired knowledge while learning new tasks. Existing methods primarily mitigate forgetting through parameter regularization or experien...

📖 Read original article


127. TransfHAR: Self-Supervised Wrist Representations for On-Demand Activity Recognition ​

Author: Aidan Bradshaw, Riku Arakawa, Xin Liu, Karan Ahuja
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15861v1 Announce Type: new Abstract: Fine-grained wrist activity recognition can support applications such as procedural step guidance and context-aware assistance, yet acquiring labeled data for every new task, user, and activity granularity remains a bottleneck. We present TransfHAR, a ...

📖 Read original article


128. Feasible and Novel Synthetic Population Generation with Tabular and Sequential Travel Attributes ​

Author: Farbod Abbasi, Zachary Patterson, Bilal Farooq
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15867v1 Announce Type: new Abstract: Synthetic populations are critical inputs for activity-based travel demand models, yet generating realistic populations from limited survey data remains challenging. Small samples miss valid attribute combinations, known as sampling zeros, and generati...

📖 Read original article


129. Deploying Frontier Agentic Technology in MOOSEnger, a Multiphysics-Capable AI Assistant ​

Author: Zaid Abulawi, Mengnan Li, Guillaume Giudicelli, Yang Liu, Cody Permann
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CE

arXiv:2608.15881v1 Announce Type: new Abstract: The Multiphysics Object-Oriented Simulation Environment (MOOSE) is an open-source finite-element framework for building multiphysics simulation applications. Using a multiphysics environment effectively demands specialized expertise, creating a barrier...

📖 Read original article


130. Layers Matter: Why Continual Learning Regularization Should Be Layer-Adaptive ​

Author: Brian B. Moser, Ahmed Anwar, Tobias Christian Nauen, Shishir Muralidhara, Federico Raue, Ren'e Schuster, Stanislav Frolov, Andreas Dengel
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15901v1 Announce Type: new Abstract: Continual learning regularizers like EWC fight forgetting by penalizing changes from previous-task parameters with per-parameter importance, typically diagonal Fisher values. Per-parameter looks more flexible than per-layer, but each layer's diagonal F...

📖 Read original article


131. Information Geometry of Message Passing ​

Author: Mykola Lukashchuk, Kyrylo Yemets, Alex Ledbetter, .{I}smail \c{S}en"oz
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.15922v1 Announce Type: new Abstract: We show that the natural-gradient stationary condition of variational inference has an edge-local form on a Forney-style factor graph. We start from the Bethe free energy and constrain a selected edge marginal to an exponential family. At a stationary ...

📖 Read original article


132. ReliaGate: Reliability Routing for Low-Stakes Wearable Stress Prediction ​

Author: Jaden Moon, Yu Wu, Arvind Pillai, Andrew Campbell
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.HC

arXiv:2608.15951v1 Announce Type: new Abstract: We study when a wearable stress system should surface a prediction rather than change it. In low-stakes reflection and summary settings, aggregate accuracy is insufficient because withholding can reduce error while leaving some people with little or no...

📖 Read original article


133. Beat the Counter First: A Baseline for Temporal-Graph Anomaly Detectors ​

Author: Omair Shafi Ahmed, Zohair Shafi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15965v1 Announce Type: new Abstract: Progress in streaming, edge-level graph anomaly detection (GAD) has been marked by increasingly elaborate architectures, from count-min-sketch chi square tests to memory-augmented attention networks. Yet the empirical gains attributable to this added c...

📖 Read original article


134. A Banach-Space Theory of Markovian Halpern Iteration for Non-Expansive Maps ​

Author: Ege C. Kaya, Arda Fazla, M. Berk Sahin, Abolfazl Hashemi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.15966v1 Announce Type: new Abstract: We study stochastic approximation of fixed points of a non-expansive operator when the oracle samples originate from a continuing Markovian trajectory. A direct block-minibatch implementation of Halpern iteration attains an expected last-iterate residu...

📖 Read original article


135. The Limits of Binding in Dual Encoders ​

Author: Kin Ian Lo
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CV

arXiv:2608.15971v1 Announce Type: new Abstract: Dual-encoder models such as CLIP score an image-caption pair by a single inner product of two independently computed unit vectors, and fail at binding, often scoring near chance when asked to distinguish "a red car and a blue dog" from "a blue car and ...

📖 Read original article


136. Fiber Fingerprints of Hidden Learning-State Dynamics ​

Author: Qinyou Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15976v1 Announce Type: new Abstract: A learning system can occupy execution states that are indistinguishable under every declared present-behavior readout yet respond differently to future training. We formalize this through fiber fingerprints: controlled future-learning response laws re...

📖 Read original article


137. Operator-Theoretic Generalization Bounds for Multitask Deep Learning ​

Author: Mahdi Mohammadigohari, Thomas Borsani, Giuseppe Di Fatta
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15982v1 Announce Type: new Abstract: We develop operator-theoretic generalization bounds for deep multi-output function classes by representing network layers as Koopman composition operators on vector-valued reproducing kernel Hilbert spaces. In vector-valued Sobolev RKHSs, we derive Rad...

📖 Read original article


138. Toward Optimal Second-Order Path-Length Guarantee for Adversarial Multi-Armed Bandits ​

Author: Mengxiao Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15996v1 Announce Type: new Abstract: We study second-order path-length regret in adversarial $K$-armed bandits against oblivious loss sequences. Bubeck et al. [2019] designed an algorithm that achieves $\widetilde{\mathcal{O}}(K+\sqrt{KQ_{\infty,1}})$ regret, where $Q_{\infty,1}$ is the f...

📖 Read original article


139. Retrieval-guided Twin Fusion with Similarity-aware Contrast for Molecule-Text Alignment ​

Author: Shunshun Gu, Shengqi Qiu, Hang Zhou, Xiao Luo
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16005v1 Announce Type: new Abstract: This paper studies the problem of molecule-text alignment, which aims to project molecules and their textual descriptions into a joint latent space for downstream tasks including molecule search and molecular property prediction. Previous approaches ty...

📖 Read original article


140. Breaking the Compression Barrier: Cross-Architecture Compression Boundary Learning via Reverse Regrowth ​

Author: Zhaocen Liu, Satvik Praveen, Yi Sheng
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.16010v1 Announce Type: new Abstract: Model compression is critical for deploying networks on resource-constrained edge devices. While pruning-based methods can significantly reduce model size, they often suffer from abrupt performance collapse beyond a sparsity thresh-old, making it diffi...

📖 Read original article


141. RagGAD: Rationale-Aware Conditional Gaussian Mixture Normalizing Flow for Unsupervised Graph Anomaly Detection ​

Author: Junxin Lu, Jing Zhao, Shiliang Sun
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16018v1 Announce Type: new Abstract: Graph anomaly detection aims to identify nodes that deviate from normal behavioral patterns within graphs. However, existing methods largely rely on the homophily assumption, which makes it difficult to distinguish spurious affinities and to capture th...

📖 Read original article


142. Group ICA 2.0: Closing the Gap Between Subjects and Group Latent Decomposition with Copula-Linked Group ICA (CoLiG-ICA) ​

Author: Oktay Agcaoglu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ME

arXiv:2608.16029v1 Announce Type: new Abstract: Group Independent Component Analysis (gICA) is widely used to decompose high-dimensional functional MRI data into interpretable brain networks. However, conventional gICA primarily identifies components shared across subjects. This group-level assumpti...

📖 Read original article


143. AdROD: HyperNetwork-based Adversarially Robust Object Detection for Autonomous Driving ​

Author: Yuting Wu, Dongfang Guo, Xiangzhong Luo, Qun Song, Rui Tan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.16031v1 Announce Type: new Abstract: Camera-based object detectors are vulnerable to physical adversarial attacks designed to suppress detections. While adversarial training and input purification offer some protection, they often overfit to specific attack distributions and fail on adapt...

📖 Read original article


144. NICE: Scale-Stable Perturbations for Graph Neural Network Explanations via Noise Corruption ​

Author: Ziluowen Luo, Jun Yin, Ruochen Liu, Ming Cheng, Shirui Pan, Chengqi Zhang, Senzhang Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16038v1 Announce Type: new Abstract: Post-hoc Graph Neural Network (GNN) explainers commonly follow a Perturb-Query paradigm, inferring the importance of graph elements based on queried predictions to perturbed inputs. However, such perturbations often introduce substantial distribution s...

📖 Read original article


145. OceanLight: Efficient Global Ocean Forecasting via Geometry-Adaptive Unstructured Mesh Representation ​

Author: Wei Wu, Xiang Wang, Hongze Leng, Qingye Min, Junxing Zhu, Junqiang Song
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16070v1 Announce Type: new Abstract: Reliable global ocean forecasting is critical for climate monitoring, marine navigation, and extreme event early warning. Physics-based ocean forecasting models impose prohibitive computational costs, while existing deep learning approaches predominant...

📖 Read original article


146. Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization ​

Author: Yixuan Wang, Yifei Chen, Haichao Zhang, Haozheng Luo, Xander Wu, Jie Ni, Yun Fu, Nuno Vasconcelos, Yijiang Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16072v1 Announce Type: new Abstract: Reinforcement learning (RL) with group-relative advantages has become the de facto standard for post-training language model reasoners. However, when optimizing multiple reward objectives, existing methods typically scalarize the reward vector with a f...

📖 Read original article


147. DeepOHeat-v2: Self-Improving Operator Learning for Fast and Trustworthy Thermal Optimization in 3D-IC Design ​

Author: Xinling Yu, Yixing Li, Ziyue Liu, Xin Ai, Zhiyu Zeng, Hai Li, Zheng Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, physics.data-an

arXiv:2608.16080v1 Announce Type: new Abstract: Thermal-aware optimization of multi-die 3D integrated circuits evaluates many designs, each a costly heat-equation solve. Operator-learning surrogates replace this solve with a fast forward pass, ideally trained from physics alone, without labeled data...

📖 Read original article


148. Towards Reasonable Molecular Structure Elucidation from Infrared Spectroscopy with Chemical Feedback ​

Author: Yusen Tan, Hongyu Zhan, Hai-tao Yu, Changxi Chi, Wenjie Du, Jun Xia
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16082v1 Announce Type: new Abstract: Infrared (IR) spectra provide characteristic signals of molecular structure, which are often interpreted by experts via functional-group identification or library matching, making the process time-consuming and ambiguous. Recent machine learning method...

📖 Read original article


149. Behaviour Is an Incomplete Measure of Reasoning Development: Cross-surface pre-arrival accessibility and the limits of developmental inference in a recurrent-depth reasoner ​

Author: Simon Lam-Muir
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16085v1 Announce Type: new Abstract: Capability development is routinely inferred from behavioural thresholds, from final checkpoints, or from what a decoder can read out of a hidden state. These quantities need not identify the same event. We study a 30M-parameter recurrent-depth relatio...

📖 Read original article


150. Unifying Graph Neural Networks Through a Common Layer Equation ​

Author: Sai Karthik Navuluru, Siddhartha Shankar Das, Bo Ni, Hongjie Chen, Yu Wang, Baris Coskunuzer, Nesreen K. Ahmed, Franck Dernoncourt, Mahantesh Halappanavar, Tyler Derr, Ryan A. Rossi, Lakshman Tamil
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16097v1 Announce Type: new Abstract: Graph neural networks are commonly described through family-specific equations whose notation obscures shared computations and structural differences. We introduce a common layer equation that represents covered architectures through seven components: ...

📖 Read original article


151. AsyTO: Asymmetric Temporal Operator for Parameter-Efficient Multivariate Time Series Forecasting ​

Author: Xiachong Lin, Du Yin, Hao Xue, Wen Hu, Imran Razzak, Arian Prabowo, Matthew Amos, Flora D. Salim
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16098v1 Announce Type: new Abstract: Multivariate time-series forecasting faces a structural dilemma: sharing one temporal predictor across variables is parameter-efficient but forces heterogeneous variables through an identical history-to-future map, whereas learning an independent predi...

📖 Read original article


152. RetroMPA: A Molecular Property-Aware Auxiliary Framework for Enhancing Retrosynthesis Prediction ​

Author: Mianzhi Liu, Fan Xiao, Zhiliang Yu, Huayang Huang, Yuke Li, Yi Yang, Wenbo Liu, Yu Wu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16111v1 Announce Type: new Abstract: Retrosynthesis is a cornerstone of drug discovery and organic synthesis. While data-driven deep learning models have shown remarkable progress, they autonomously learn reaction patterns from extensive datasets with limited integration of established ch...

📖 Read original article


153. Multi-Feature Riemannian Hypergraph for Online Test-Time Adaptation of Motor Imagery Brain-Computer Interface ​

Author: Siqi Li (Peking University, Chinese Institute for Brain Research, Beijing), Zhi Li (NeuCyber Neurotech), Tong Liu (NeuCyber Neurotech), Shuai Zhang (NeuCyber Neurotech), Yanfei Jia (Beijing Medical University), Zhiqiang Yi (Beijing Medical University), Jue Xie (NeuCyber Neurotech), Ni Ji (Chinese Academy of Medical Sciences & Peking Union Medical College, Chinese Institute for Brain Research, Beijing)
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, eess.SP, q-bio.NC

arXiv:2608.16134v1 Announce Type: new Abstract: In clinical motor imagery brain-computer interface (MI-BCI) decoding, cross-day transferability and online operation remain two critical challenges. Hypergraphs can improve transferability by capturing higher-order sample relationships, yet existing hy...

📖 Read original article


154. REFLEX: Reflexive Equilibrium Fixed-point Learning for Endogenous eXchanges ​

Author: Vignesh Nagarajan, Shriraghav Ashok
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.GT

arXiv:2608.16155v1 Announce Type: new Abstract: In over-the-counter corporate bond markets, dealers compete for client trades by quoting bid and ask prices. Tighter quotes attract more business, but also informed customers more likely to trade ahead of adverse price moves, leaving the dealer holding...

📖 Read original article


155. A Tree-Structured Approach for Phishing Template and Attacker Attribution Analysis ​

Author: Unai Agirre, Imanol Jerico, Felipe Casta~no, Andrea Venturi, Francesco Zola
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET

arXiv:2608.16158v1 Announce Type: new Abstract: Phishing remains a persistent and evolving cybersecurity threat, with attack volumes reaching record levels. This growth is driven by the industrialization of phishing through widely available phishing kits and reusable templates, which enable cybercri...

📖 Read original article


156. Demystifying Oversmoothing in Sheaf Neural Networks: An Index-Theoretic Criterion ​

Author: Junwen Dong, Yuhan Peng, Hao Li, Huitao Feng, Kelin Xia
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math.DG

arXiv:2608.16180v1 Announce Type: new Abstract: To combat oversmoothing in Graph Convolutional Networks, Sheaf Neural Networks (SNNs) were proposed as a generalization by equipping the graph with a sheaf structure and replacing the graph Laplacian with a sheaf Laplacian $\mathcal{L}$. Existing analy...

📖 Read original article


157. Understanding and Stabilizing Deep Q-Learning via Controlled Bootstrapping and Regulated Value Dynamics ​

Author: Bozhou Chen, Yongyi Wang, Hanyu Liu, Xionghui Yang, Wenxin Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16182v1 Announce Type: new Abstract: Deep Q-learning (DQL) has achieved remarkable empirical success in reinforcement learning, yet its training process remains notoriously unstable. Existing studies often attribute instability to isolated factors such as overestimation bias or representa...

📖 Read original article


158. Multi-Granularity Sentiment Integration for LLM-Based Multimodal Sentiment Analysis ​

Author: Shanshan Lin, Yuesheng Wu, Chao Chen, Yizhe Yang, Zhihao Chen, Zexian Yang, Xiangwen Liao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16201v1 Announce Type: new Abstract: Multimodal sentiment analysis (MSA) aims to predict sentiment polarity and intensity from heterogeneous inputs such as text, audio, and vision. While large language models (LLMs) offer strong semantic priors for MSA, effectively incorporating audio and...

📖 Read original article


159. Conditional Evaluation of Language Models with Cheap Auxiliary Signals ​

Author: Zhi Zhang, Lingfeng Lyu, Yue Kang, Doudou Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.16210v1 Announce Type: new Abstract: Aggregate accuracy hides where models succeed and fail. Estimating conditional performance profiles from gold labels alone is expensive, while cheap auxiliary signals such as LLM-judge scores, pairwise comparisons, confidence scores, and judge-disagree...

📖 Read original article


160. Quantifying the Gap Between Laboratory Battery Test Patterns and Field Duty Profiles ​

Author: Chunyang Zhao, Chresten Tr{\ae}holt
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16212v1 Announce Type: new Abstract: Laboratory battery tests provide the main empirical basis for battery performance and degradation studies, but their operating patterns do not directly represent field duty profiles. This paper quantifies the gap by comparing six accessible evidence so...

📖 Read original article


161. Beyond Peak Backlog: Conditional Energy and Temporal Geometry in Capacity-Constrained Delayed Bandit Optimization ​

Author: Anling Xiang, Yuwen Yang, Yang Shen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16216v1 Announce Type: new Abstract: What is the right delay complexity when a learner can track only $C$ pending feedback items and discarded feedback is permanently lost? Existing one-point bandit convex optimization guarantees in this model pay $\sqrt{T\sigma_{\max}}$, where $\sigma_{...

📖 Read original article


162. A Privacy Study of Sparse Collaborative Inference ​

Author: Maximilian Andreas Hoefler, Karsten Mueller, Wojciech Samek
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16236v1 Announce Type: new Abstract: Collaborative inference (CI) splits a model between an edge device and a server, whereby the client computes an intermediate activation, transmits it, and the server completes the computation. This raises two concerns, the communication cost of the tra...

📖 Read original article


163. Optimizing Multi-Market Participation of Battery and Electrolyser Systems Based on Field Performance ​

Author: Chunyang Zhao, Stoyan Trenchev, Shi You, Chresten Tr{\ae}holt
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16238v1 Announce Type: new Abstract: The increasing share of renewable energy in power systems creates a need for fast-response and flexible resources to maintain system stability. With the expansion of electricity markets and ancillary service products, opportunities arise to stack reven...

📖 Read original article


164. The Trade-off Between Covariate Dependence and Latent Structure in Representation Learning ​

Author: Ma{\l}gorzata {\L}az\k{e}cka, Ewa Szczurek
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16245v1 Announce Type: new Abstract: Disentangled representation learning seeks latent representations whose indicidual dimensions each align with a distinct covariate. Unsupervised approaches typically target latent dimension independence, yet this gives no guarantee that the resulting d...

📖 Read original article


165. SAUL: Sharpness-Aware Augmented-Lagrangian Unlearning ​

Author: Jaewan Choi, Junyoung Yang, Sangdon Park
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16249v1 Announce Type: new Abstract: Machine unlearning in Large Language Models (LLMs) faces a critical trade-off between erasing target knowledge and preserving general utility. We propose SAUL (Sharpness-Aware Augmented-Lagrangian Unlearning), which formulates unlearning as a constrain...

📖 Read original article


166. Efficient Coreset Selection via K-Nearest Neighbor Graphs ​

Author: Yingfan Liu, Leiyu Zhang, Jiadong Xie, Mingzhe Wang, Jeffrey Xu Yu, Jiangtao Cui
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16270v1 Announce Type: new Abstract: Coreset selection reduces the cost of model training by replacing a large training set with a small representative subset. Existing gradient-approximation coreset methods such as CRAIG and cluster-based variants can preserve model accuracy. Still, thei...

📖 Read original article


167. Foresight-England: Development of a National-Scale Generative AI Model of Electronic Health Records for Medical Event Prediction across the COVID-19 Pandemic ​

Author: Simon Ellershaw, Christopher Tomlinson, Zeljko Kraljevic, Spiros Denaxas, Harry Hemingway, Cathie Sudlow, Angela M. Wood, Anoop D. Shah, Richard Dobson
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16273v1 Announce Type: new Abstract: Foresight-England (Foresight-E) is the first national-scale generative foundation model of electronic health records (EHRs), developed as a research pilot strictly for COVID-19 research. We evaluated its ability to model the direct and indirect effects...

📖 Read original article


168. SCALE: State-Calibrated Latent Embeddings for JEPA Planning in the Right Geometry ​

Author: Jiaming Hu, Yan Zheng, Tian Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16287v1 Announce Type: new Abstract: Joint-embedding predictive world models plan by scoring predicted terminal embeddings against a goal embedding using a cost defined on the representation itself. Two prominent strategies for obtaining non-collapsed representations are to inherit a pret...

📖 Read original article


169. Advancing Open and Reproducible Relational Learning: RelArena-$\alpha$, TabPFN-Rel and RPI ​

Author: Adrian Hayler, Klemens Fl"oge, Alan Arazi, Rishabh Ranjan, Jure Leskovec, Felix Birkel, Brendan Roof, Anurag Garg, Kristina Collins, Lydia Sidhoum, Jonas K"ubler, Siyuan Guo, Oscar Key, Jan Hendrik Metzen, Rylee Grace, David Salinas, Arthur Cahu, Simon Bing, Benjamin J"ager, Tuana \c{C}elik, Mihir Manium, Vitor Monteiro, Jake Robertson, Jerry Chen, Eliott Kalfon, Tom'as Pereda, Lilly Wehrhahn, Dominik Safaric, Tobias Schroeder, Georg Grab, Diana Kriuchkova, Clara Cornu, Philipp Singer, Nick Erickson, Vahid Balazadeh, Marie Salmon, Simone Alessi, K"ur\c{s}at Kaya, Philipp Jund, L'eo Grinsztajn, Yann LeCun, Bernhard Sch"olkopf, Madelon Hulsebos, Lennart Purucker, Sauraj Gambhir, Frank Hutter, Noah Hollmann
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16319v1 Announce Type: new Abstract: This first release of Prior Labs in relational learning shows our continued commitment to open science. We open-source three pieces of software that we expect to accelerate research in the field towards meaningful real-world impact. We aim to steer fur...

📖 Read original article


170. Transfer Learning of Keystroke Dynamics for Cross-Device User Authentication ​

Author: Nuwan Kaluarachchi, Sevvandi Kandanaarachchi, Kristen Moore, Arathi Arakala, Conrad Sanderson
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.HC

arXiv:2608.16334v1 Announce Type: new Abstract: Keystroke dynamics (typing patterns) can be used as a behavioural biometric modality for user authentication, with applications such as fraud prevention. While the modality has been shown to work well for single device authentication, its application t...

📖 Read original article


171. Task-Anchored Representation Shaping for Pre-Trained Model-Based Continual Learning ​

Author: Zhiming Xu, Huiyu Yi, Zhen-Hao Xie, Baile Xu, Furao Shen, Jian Zhao, Suorong Yang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16345v1 Announce Type: new Abstract: Pre-trained models (PTMs) provide a strong foundation for continual learning by offering stable representations that facilitate lightweight adaptation to new tasks. However, adapting well to each task does not ensure reliable inference over all learned...

📖 Read original article


172. OceanDepths: A Global Dataset of Paired Subsurface and Surface Ocean Observations ​

Author: Simon Donike, Ruben Cartuyvels, Antonino Ian Ferola, Elisa Carli, Diego Fernandez Prieto, Marie-Helene Rio
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.16373v1 Announce Type: new Abstract: Despite comprising over 70% of its surface, the world's oceans are critically underobserved compared to the land surface or the atmosphere.Understanding the global ocean requires jointly observing its surface and subsurface structure, yet no standardi...

📖 Read original article


173. Coverage-Maximizing Multinomial Subset Routing under Operational Constraints ​

Author: Quan Zhou, Yiyan Huang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16375v1 Announce Type: new Abstract: We introduce Multinomial Subset Routing (MSR), a new online routing framework over $K$ experts in which the learner keeps a multinomial routing policy instead of a deterministic subset of experts. At each round, the learner samples $M$ experts i.i.d. f...

📖 Read original article


174. FETERS: Few-Shot Early Time-Series Classification via Effective Ratio Selection ​

Author: Chen-An Tai, Yujia Wu, Vincent S. Tseng
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16385v1 Announce Type: new Abstract: Early time-series classification (ETSC) aims to make accurate predictions from partially observed time series as early as possible. Although various stopping mechanisms and feature learning strategies have been developed for ETSC, most existing methods...

📖 Read original article


175. SoftModel: A Neural Model That Grows Its Own Topology -- Governed Structural Growth for Continual In-Service Learning ​

Author: Zhoumin Xie
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16409v1 Announce Type: new Abstract: Today, a neural system is almost always used in two phases -- trained, then deployed -- and in that regime it freezes twice: training ends, and the topology itself was never a degree of freedom. We take the opposite premise as an axiom -- total plastic...

📖 Read original article


176. TRACE-CASH: Trial-History-Conditioned Reinforcement Learning for Adaptive Configuration Exploration in Time-Series CASH ​

Author: Yu-Han Huang, Yujia Wu, Vincent S. Tseng
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16410v1 Announce Type: new Abstract: Combined algorithm selection and hyperparameter optimization (CASH) searches a conditional space in which the selected model determines which hyperparameters are active. In time-series forecasting, temporal choices, chronological validation, and costly...

📖 Read original article


177. Evolving Executable Pipeline Programs for AutoML with Language Models ​

Author: Sofoklis Kitharidis, Cor J. Veenman, Jan N. van Rijn, Thomas B"ack, Niki van Stein
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2608.16416v1 Announce Type: new Abstract: Automated machine learning (AutoML) systems search for pipelines within a space of preprocessing operators, learners, and hyper-parameters specified in advance: they can select and tune known components, but cannot produce structure outside that space....

📖 Read original article


178. PertMind: Eliciting Emergent Biological Reasoning in LLM via Reinforcement Learning on Cellular Perturbation Data ​

Author: Zhenchao Tang, Xiaogang Xu, Tianxu Lv, Jiahui Guan, Jiale Zhou, Haohuai He, Zhi Song, Hanbo Huang, Jiehui Huang, Jiafei Wu, Zhe Liu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.QM

arXiv:2608.16419v1 Announce Type: new Abstract: Large language models can describe mechanisms, yet scalable post-training still depends on costly, manually curated biological reasoning traces. Here we show that cellular perturbation atlases can instead become reinforcement-learning environments, whe...

📖 Read original article


179. Localized TabICLv2: Scaling Tabular In-Context Learning through k-NN ​

Author: Beimnet Bekele Guta
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16429v1 Announce Type: new Abstract: Foundational models for tabular data have made significant progress in recent years, with TabICLv2 reporting state-of-the-art performance on several tabular classification tasks. However, full-context tabular ICL still suffers from attention cost that ...

📖 Read original article


180. Reference-free logged energy-oracle recovery for neural approximations of symmetric coercive variational problems: conforming Riesz reconstruction and archive-level selection ​

Author: Karim Bounja, Lahcen Laayouni, Boujemaa Achchab, Abdeljalil Sakat
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2608.16473v1 Announce Type: new Abstract: Neural PDE training yields a finite checkpoint archive, yet its logged energy errors are inaccessible without the exact solution, while loss-based selection does not necessarily recover the logged energy oracle. For admissible neural approximations of ...

📖 Read original article


181. Pallas: A Proactive KV Cache Migration Framework for LLM Inference in AI-RAN ​

Author: Tianhang Ding, Jianchun Liu, Hongli Xu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16477v1 Announce Type: new Abstract: AI-RAN brings large language model (LLM) serving close to mobile users, but cellular handover can separate an active request from its inference state: the user attaches to a target base station (gNB) while the large and growing key-value (KV) cache rem...

📖 Read original article


182. Graph Machine Learning: An Opportunity for Power Systems ​

Author: Martin Sadric, Sebastian P"utz, Christian Nauck, Veit Hagenmeyer, Frank Hellmann, Dirk Witthaut, Benjamin Sch"afer
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE, cs.SY, eess.SY

arXiv:2608.16494v1 Announce Type: new Abstract: Modern power systems face growing operational complexity driven by the integration of renewable energy sources, decentralization, and the need for real-time decision-making across a wide range of timescales. Addressing these challenges traditionally re...

📖 Read original article


183. When Tool-Backed Skill Retrieval Fails: Source-Style Collapse in Executable Capability Retrieval ​

Author: Yiqi Liu, Joseph James, Yang Wang, Chenghao Xiao, Chenghua Lin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2608.16502v1 Announce Type: new Abstract: Large-scale agents increasingly rely on retrieval to access external capabilities. We study this retrieval gate in structured tools and APIs, a measurable class of tool-backed executable skills that must be surfaced before an agent can plan, incorporat...

📖 Read original article


184. One Residual with Three Reuses: A Wristband Front End for Gesture Sensing ​

Author: Sam Rifaki
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.HC

arXiv:2608.16542v1 Announce Type: new Abstract: Continuous wrist-worn hand sensing for gesture interfaces and motor symptom monitoring needs an always-on front end that fits inside a coin-cell power budget while pairing a micro-electro-mechanical-systems (MEMS) inertial measurement unit (IMU) with a...

📖 Read original article


185. Learning Generalizable Reconstruction of High-Dimensional Neural Dynamics ​

Author: Anima Kujur, Zahra Monfared
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16569v1 Announce Type: new Abstract: Accurate reconstruction of long-duration neural recordings is challenging because local field potentials (LFPs) are high-resolution, multichannel, transient, and variable across subjects. We present PCA-DMD, a scalable operator-theoretic framework that...

📖 Read original article


186. Variational Outlier-Robust Gaussian Process Regression with Generative Modeling ​

Author: Arslan Majal, Aamir Hussain Chughtai
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16606v1 Announce Type: new Abstract: Outliers can substantially distort Gaussian process regression (GPR) due to its conventional Gaussian observation likelihood, leading to inaccurate model learning and prediction. To address this limitation, this article introduces a generative GPR mode...

📖 Read original article


187. Hoeffding adaptive splitting trees for data stream classification with concept drift and ensemble learning ​

Author: Daniel Nowak Assis, Jean Paul Barddal, Fabr'icio Enembreck
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16659v1 Announce Type: new Abstract: Ensembles of decision trees are well-established methods for data stream classification. In ensemble learning, Hoeffding Trees are widely adopted as base learners, performing periodic split attempts according to the Hoeffding bound. Recent studies, how...

📖 Read original article


188. UniTAC: Universal Task-Aware Compression via Weighted Distortion Measures ​

Author: Homa Esfahanizadeh, Matin Mortaheb, Jinfeng Du, Harish Viswanathan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, cs.MM, math.IT

arXiv:2608.16696v1 Announce Type: new Abstract: Physical AI systems such as autonomous vehicles and robots rely on timely exchange of high-dimensional sensory signals under tight bandwidth, latency, and energy budgets. Because the task driving downstream decisions evolves over time, a task-specific ...

📖 Read original article


189. Learning to Unlearn: Machine Unlearning via Learning the Unlearning Behaviors ​

Author: Hang Zhang, Kaifeng Zhang, Yixiao Ma, Weijie Xu, Ye Zhu, Kai Ming Ting
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16700v1 Announce Type: new Abstract: Various machine unlearning techniques have been developed in response to privacy legislation requirements, enabling individuals to exercise their legal right to have their data $D_f$ removed from a machine learning model. This process is typically acco...

📖 Read original article


190. The Ethical Decision Head: Operationalizing Normative Ethics in Autonomous Vehicles via Reinforcement Learning from Human Feedback ​

Author: Thomas Mbrice, Ammar Ali, Sami Mian, Khai Hern Low, Eric Chen, Arshia Aghajani, Wolf Sch"afer, Amin Shirangi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16710v1 Announce Type: new Abstract: As autonomous vehicles (AVs) approach Level 4 and Level 5 operational capability [SAE International, 2018], their on- board decision systems must handle not only safety-critical locomotion but also their subsequent moral weight. This paper details the ...

📖 Read original article


191. Le Critique: Privileged Value Functions for LLM Reinforcement Learning ​

Author: Siddarth Venkatraman, Matthieu Dinot, Laurence Aitchison
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16739v1 Announce Type: new Abstract: Reinforcement learning algorithms for Large Language Models (LLMs) are largely distinguished by their variance reduction strategy. Group-relative methods like GRPO reduce gradient variance by sampling multiple rollouts per prompt, but provide only sequ...

📖 Read original article


192. Would this change your answer? Evaluating Explanations of LLM Behavior In The Wild with Counterfactual Experiments ​

Author: Adam Karvonen, Euan Ong, Subhash Kantamneni, Samuel Marks
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16747v1 Announce Type: new Abstract: Many areas of AI research, such as language model interpretability and chain of thought faithfulness, seek to explain model behaviors. But what constitutes a "good" explanation? In this work, we evaluate explanations through the lens of counterfactual ...

📖 Read original article


193. On the Principles Behind Neural Network Optimizers ​

Author: Yushun Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2608.16760v1 Announce Type: new Abstract: Reliable optimization is central to neural network (NN) training, yet Adam, the default optimizer for modern LLMs, rests on a fragile foundation. This thesis develops a principled grounding for Adam and motivates new designs. First, we revisit Adam's d...

📖 Read original article


194. Beyond $L_2$: Generalizing Abductive Latent Explanations to Diverse Prototype-Based Architectures ​

Author: Jules Soria, Alban Grastien, Romain Xu-Darme, Julien Girard-Satabin, Zakaria Chihani, Daniela Cancila
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16773v1 Announce Type: new Abstract: Prototype-based neural networks are hailed as interpretable-by-design architectures. Recently, Abductive Latent Explanations (ALE) were introduced to provide formal, mathematically guaranteed explanations that leverage the intrinsic structure of these ...

📖 Read original article


195. GEO-Flag: Detecting and Measuring GEO-Optimized Web Content ​

Author: Junjie Chu, Ye Leng, Mingjie Li, Yun Shen, Xinyue Shen, Yang Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.IR

arXiv:2608.16824v1 Announce Type: new Abstract: Generative Engine Optimization (GEO) modifies web content to increase its likelihood of being selected and cited by generative search engines. This can give strategically optimized pages visibility disproportionate to their authority or relevance and e...

📖 Read original article


196. CaliBench: Are the Stochastic Dynamics of Video World Models Physically Calibrated? ​

Author: Jonathan Sadeghi, Jenny Seidenschwarz, Jesse Allardice, Sirish Srinivasan, Benjamin Graham, Jeffrey Hawke
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.16829v1 Announce Type: new Abstract: Video world models approximate the stochastic distribution of physical outcomes through generative sampling, but existing benchmarks score individual generations or compare distributions coarsely over a whole dataset, leaving the fine-grained aleatoric...

📖 Read original article


197. Time-Aware Validation of Machine Learning Fuel Consumption Models: Evidence from 1\,Hz Operational Data, CCGS \textit ​

Author: Samarasimha Reddy Chittamuru, Ayhan Akinturk, Allison Kennedy, Joshua Barnes, Matthew Hamilton
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16833v1 Announce Type: new Abstract: Ship fuel consumption (SFC) prediction supports vessel operation optimisation, emissions estimation, and decision support systems (DSS) for sustainable maritime transportation. Numerous data-driven fuel models have been developed over the past two deca...

📖 Read original article


198. Proteus: Incremental Memory Activation for Long-Context Sequence Modeling ​

Author: Reza Bayat, Ali Behrouz, Vahab Mirrokni, Aaron Courville
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.16844v1 Announce Type: new Abstract: The quadratic cost of attention-based sequence models for long contexts has motivated a growing line of research on memory-based models that can compress context into a compact state. However, most existing memory models expose a static memory througho...

📖 Read original article


199. Data-Efficient and Interpretable Classification of Circulating Tumor Cell Phenotypes in Microfluidic Devices via Deep Learning ​

Author: Serena Su, Yifan Wang, Senwei Liang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16870v1 Announce Type: new Abstract: Accurate classification of circulating tumor cell (CTC) phenotypes can provide valuable information for assessing metastatic potential. Label free microfluidic devices provide a hydrodynamic obstacle course that transforms subtle biophysical characteri...

📖 Read original article


200. An Analytical-Prior Framework for Data-Efficient Prediction of Sound-Reduction Frequencies in Rectangular Side-Branch Helmholtz Resonators ​

Author: Jiaming Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16873v1 Announce Type: new Abstract: High-fidelity finite-element simulations can provide accurate numerical predictions for side-branch resonators, but large simulation datasets are expensive to generate and purely data-driven surrogates may become unreliable when simulation-labelled dat...

📖 Read original article


201. Q-based Variational Inverse Reinforcement Learning ​

Author: Ondrej Bajgar, Peter Tisnikar, Alessandro Abate, Konstantinos Gatsis, Maike Osborne
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.16888v1 Announce Type: new Abstract: The development of safe and beneficial AI requires that systems can learn and act in accordance with human preferences. However, explicitly specifying these preferences by hand is often infeasible. Inverse reinforcement learning (IRL) addresses this ch...

📖 Read original article


202. Auxiliary uncertainty signals for LLM-assisted systematic review screening: a benchmark across eight Cohen drug-class reviews ​

Author: Arya Rahgozar, Pouria Mortezaagha
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.DL, cs.IR, cs.LG

arXiv:2608.14551v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for title-abstract screening in systematic reviews, but their decisions lack calibrated uncertainty. We show that an auxiliary BERT+GCN classifier supplies a structured uncertainty signal that improv...

📖 Read original article


203. When to Communicate: Belief Distributions and KL Divergence for Principled Gating in Multi-Agent RL ​

Author: Teoman Kaman
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.14559v1 Announce Type: cross Abstract: Effective communication in multi-agent reinforcement learning requires agents to decide not only \textit{what} to communicate, but when? Existing approaches either communicate at every timestep or learn a binary gate through REINFORCE policy gradient...

📖 Read original article


204. Longitudinal and Graph-Augmented Prediction of Adolescent Substance Use Onset in the ABCD Study ​

Author: Yixuan He, Jinni Su, Yun Kang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG, stat.AP

arXiv:2608.14578v1 Announce Type: cross Abstract: Early identification of adolescent substance-use risk is an important prevention challenge, yet the relative value of baseline characteristics, longitudinal trajectories, and relational context remains unclear. Using data from approximately 11,860 pa...

📖 Read original article


205. Characterization of Thermal Systems from Noisy and Low-resolution Measurements Using Dynamic Mode Decomposition ​

Author: M. E. P. Silva, L. S. Araujo, F. T. Colombo, A. Cunha Jr, S. da Silva
Published: 8/18/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG, eess.SP, physics.class-ph

arXiv:2608.14581v1 Announce Type: cross Abstract: Thermal monitoring in practical applications is often constrained by sparse sensing, measurement noise, and limited spatial resolution, which hinder the identification of heat transfer dynamics. In such settings, calibrating high-fidelity physical mo...

📖 Read original article


206. Evaluating the impact of adversarial traffic patterns on vanet communication using veins simulation ​

Author: Henry Agyapong
Published: 8/18/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2608.14583v1 Announce Type: cross Abstract: Vehicular Ad Hoc Networks (VANETs) are a key component of intelligent transportation systems, enabling real-time communication between vehicles. However, their open and dynamic nature makes them highly vulnerable to adversarial behaviors that can dis...

📖 Read original article


207. 6G Native AI and Channel Foundation Models ​

Author: Shugong Xu, Jun Jiang, Yuan Gao
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.IT, cs.LG, math.IT

arXiv:2608.14591v1 Announce Type: cross Abstract: The integration of artificial intelligence (AI) and wireless communications is widely regarded as a core objective of sixth-generation (6G) systems. However, both the meaning of native AI and the type of AI capability that should be embedded into fut...

📖 Read original article


208. Demo: Real-time Generative Multicasting with On-Device Intent-aware Semantic Decomposition ​

Author: Xinkai Liu, Mahdi Boloursaz Mashhadi, Yi Ma, Rahim Tafazolli
Published: 8/18/2026, 4:00:00 AM
Categories: cs.NI, cs.LG, cs.MM

arXiv:2608.14600v1 Announce Type: cross Abstract: We present a demonstration for generative multicasting with on-device, intent-aware semantic decomposition. At the transmitter, DNN-based segmentation extracts a semantic map from the source video, decomposing it into multiple sub-signal classes base...

📖 Read original article


209. LLM Safety Alignment in Low-Resource Languages: A Systematic Literature Review ​

Author: Valdini Douglace Lemofouet, Blessing Ngozi Uzor, Paula Chikaodinaka Anyanwu, Danielle Blanche Kapsa, Sukairaj Hafiz Imam, P Sam Sahil, Abigail Oppong, Tassallah Abdullahi, Clemencia Siro, Idris Abdulmumin, Seid Muhie Yimam, Shamsuddeen Hassan Muhammad
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.14626v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved substantial progress in safety alignment, yet their safety guarantees remain significantly weaker in low-resource and multilingual settings than in high-resource languages. In this paper, we conduct a System...

📖 Read original article


210. Wolff-Parkinson-White Detection at 471:1 Class Imbalance: A Leakage-Controlled Study of the Data Bottleneck ​

Author: Nathael Altman
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.14633v1 Announce Type: cross Abstract: Wolff-Parkinson-White (WPW) syndrome is a congenital cardiac pre-excitation, clinically important and often missed on the resting 12-lead ECG. Detection is hard: the signature is subtle and the condition rare. We pool two public 12-lead corpora, PTB-...

📖 Read original article


211. Belayer: Efficient Fault Tolerance for LLM Agentic RL Training ​

Author: Jiecheng Zhou, Qinghao Hu, Peng Sun, Xingcheng Zhang, Weiming Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.14635v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly trained with reinforcement learning in long-horizon, sandboxed environments. Unlike conventional RL, agentic RL couples GPU-intensive rollout engines with stateful environment containers whose action...

📖 Read original article


212. Stop Indexing at Full Precision: Revisiting Clustering for Vector Embeddings ​

Author: Leonardo Kuffo, Peter Boncz
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DB, cs.AI, cs.LG

arXiv:2608.14648v1 Announce Type: cross Abstract: In this study, we revisit three widely used techniques in vector search and utilize them to optimize vector embedding indexing through clustering: dimensionality reduction, quantization, and dimension pruning. We propose an indexing pipeline in which...

📖 Read original article


213. When Uncertainty Isn't Enough: An Empirical Study of Self-Correction in Code Generation ​

Author: Pranav Rakasi, Maanas Lalwani, Arnav Srivastava, Arya Palanivel, Tinuade Adeleke, Ruizhe Li, Sean Wu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE

arXiv:2608.14659v1 Announce Type: cross Abstract: Large language models for code generation often produce incorrect solutions without reliable indicators of failure. We study whether uncertainty estimation methods developed for natural language transfer to code generation, and whether such signals c...

📖 Read original article


214. Does the Heart Show Your Pain? Tackling the X-ITE Pain Challenge with Self-Supervised ECG Representation Learning ​

Author: Dominika Kunc, Przemys{\l}aw Kazienko, Stanis{\l}aw Saganowski
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG

arXiv:2608.14662v1 Announce Type: cross Abstract: Accurate recognition of pain using physiological signals remains a challenging problem due to pain's subjective nature and high inter-individual variability. In this study, we investigate self-supervised representation learning (SSL) methods applied ...

📖 Read original article


215. Phase-Aware CNN for Real-Time 5G/6G Channel Estimation with Hardware-in-the-loop Validation ​

Author: Javad Zolfaghari-Bengar, Rakibul Rony, Elisa Gomez-de-Lope, Alejandro Villena-Rodriguez, Abhinav Mahadevan, Nicolas Kourtellis
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.14676v1 Announce Type: cross Abstract: In 5G/6G wireless systems, accurate and timely channel estimation is critical to ensure reliable communication under complex, fast-changing radio conditions. This work focuses on pilot-based channel estimation using deep learning to reconstruct both ...

📖 Read original article


216. Offline Ambient-Controlled Latent Diffusion: Architecture, Telemetry, and On-Device Evaluation ​

Author: Lech Kalinowski, Artur Morys-Magiera, Piotr Mi{\l}kowski
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG

arXiv:2608.14677v1 Announce Type: cross Abstract: Most mobile image-generation applications are thin clients over cloud services, leaving outputs hard to audit. We present an Android latent-diffusion application that runs entirely on-device and is driven by the ambient-light sensor rather than a tex...

📖 Read original article


217. A Low-Cost IoT Device for Environmental Monitoring and Embedded Solar Forecasting with On-Device Incremental Learning ​

Author: Erick Michel Lara Pinal, Abhinav Das, Stephan Schl"uter
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.14698v1 Announce Type: cross Abstract: Hyperlocal meteorological sensing is essential for accurate solar photovoltaic forecasting, yet professional-grade meteorological stations require investments easily exceeding 1000~USD per node, making distributed deployments economically inaccessibl...

📖 Read original article


218. Deep Analog: Open-Set Film Emulation with Reference-Conditioned 3D LUTs ​

Author: Yitong Mu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.LG, cs.MM

arXiv:2608.14702v1 Announce Type: cross Abstract: Film emulation reproduces the look of an analog film stock on a new digital photograph. We target its open-set form -- matching any reference film frame from a single example -- with a 3D lookup table (LUT) predicted from that reference. Real-time im...

📖 Read original article


219. On Cross-Validation for Hyperparameter Optimization of Deep Learning Image Classifiers ​

Author: Ljubomir Buturovic (East Palo Alto, United States)
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.14705v1 Announce Type: cross Abstract: Hyperparameter optimization (HPO) can materially affect the performance of deep learning (DL) image classifiers, but there is little empirical guidance on how to derive the validation signal that drives it, especially for the small sample sizes commo...

📖 Read original article


220. Equilibrium Forcing: Adaptive Video Generation Without Noise Conditioning ​

Author: Hansen Jin Lillemark, Alex Rojas, Zachary Novack, Runqian Wang, Yilun Du, Yian Ma, Taylor Berg-Kirkpatrick, Rose Yu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.14706v1 Announce Type: cross Abstract: Standard autoregressive video generation algorithms based on Diffusion and Flow Matching rely on rigid training objectives and static sampling schedules, limiting inference procedures from adapting to the data. We introduce Equilibrium Forcing (EqF),...

📖 Read original article


221. Hardware-in-the-Loop Phase-Aware CNN for Real-Time 5G Channel Estimation ​

Author: Javad Zolfaghari-Bengar, Rakibul Rony, Elisa Gomez-de-Lope, Alejandro Villena-Rodriguez, Abhinav Mahadevan, Nicolas Kourtellis
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.CV, cs.IT, cs.LG, math.IT

arXiv:2608.14709v1 Announce Type: cross Abstract: This demo presents real-time AI-based uplink channel-estimation inference using data collected from a hardware-in-the-loop 5G platform. The data-collection setup integrates commercial RF signal generation, programmable channel emulation, an O-RAN Rad...

📖 Read original article


222. Which Question Is Your Attention Metric Answering? Attention Rows as Compositional Data ​

Author: Marios Papamichalis, Regina Ruane
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, math.ST, stat.TH

arXiv:2608.14712v1 Announce Type: cross Abstract: Each row of a transformer's attention matrix is a probability distribution over tokens, and in trained models most of that probability lands on a single \emph{sink} token, usually the first. Standard tools for comparing attention rows (cosine similar...

📖 Read original article


223. Koopman early warning signals for bifurcation and rate-induced tipping ​

Author: Juan Nathaniel, Carla Roesch, Derek DeSantis, Parvathi Kooloth, Hang Fan, Valerio Lucarini, Anastasia Romanou, Pierre Gentine
Published: 8/18/2026, 4:00:00 AM
Categories: nlin.CD, cs.LG

arXiv:2608.14716v1 Announce Type: cross Abstract: Abrupt transitions in complex systems are often preceded by early warning signals. However, most indicators rely on the notion of critical slowing down and do not generally extend to rate-induced tipping where transitions can occur without local loss...

📖 Read original article


224. Local Gains and Fixed-Assignment Set Losses in Shared Set Decoders ​

Author: Ze Zhang, Yang Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.14717v1 Announce Type: cross Abstract: A query-relation deletion can improve the edited slot while reducing the utility of the prediction set that contains it. We study this tension in two related ResNet-50 DETR-family checkpoints using recorded, selection-conditional evidence from 710 pa...

📖 Read original article


225. A Vision Transformer for ECG-Based Detection of Left Ventricular Systolic Dysfunction Across Multiple Clinical Sites ​

Author: Burcu Ozek, Aruna Mohan, David Vorchheimer, Daniel Weiss, Eyal Kedar, Tamar Sobol, Or Zilbershot, Fatemeh Afghah
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.14723v1 Announce Type: cross Abstract: Reduced left ventricular ejection fraction (LVEF) is frequently asymptomatic and often detected only after advanced heart failure develops. Electrocardiograms are recorded routinely yet underused for this condition, because reduced LVEF has no single...

📖 Read original article


226. Privacy-Preserving Dataset Curation for Kuala Lumpur Urban Traffic: Grounded Vision-Language Detection with Spatial Vehicle-Context Filtering ​

Author: Mohammed Abdul Al Arafat Tanzin, Rudzidatul Akmam Dziyauddin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.14724v1 Announce Type: cross Abstract: The rapid advancement of intelligent transportation systems and autonomous driving relies heavily on multi-modal urban traffic datasets. However, curating high-fidelity video imagery in complex tropical urban environments---specifically Kuala Lumpur,...

📖 Read original article


227. IP Protection in the Era of Visual Generative AI: A Survey ​

Author: Zhuan Shi, Shunchang Liu, Alireza Dehghanpour Farashah, Qian Yang, Han Yu, Cao Yang, Chaochao Chen, Yuping Yan, Yaochu Jin, Golnoosh Farnadi, Lingjuan Lyu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.CR, cs.LG

arXiv:2608.14730v1 Announce Type: cross Abstract: The rapid evolution of visual generative AI has introduced a wide range of intellectual property risks, spanning the unauthorized learning, reproduction, extraction, misuse, and redistribution of protected data and model assets. To address these risk...

📖 Read original article


228. PandasCorpus: A Resource of Real-World Pandas Workflows and Usage Patterns ​

Author: Syrym Abdikhan, Mazhar Hameed
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2608.14742v1 Announce Type: cross Abstract: Pandas has emerged as the de facto library for data processing and machine learning, widely used for tasks, such as data loading, transformation, and analysis. Despite its ubiquity, there has been limited systematic investigation into how Pandas is u...

📖 Read original article


229. The Note-Chord-Voice Framework: Structured Source Separation and Causal Inference for EV Charging Data ​

Author: Jiajie Chen, Jinfeng Li
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.LG, cs.SD

arXiv:2608.14756v1 Announce Type: cross Abstract: Real-world EV charging data exhibit three interlocking pathologies: hardware fragmentation (network timeouts and billing resets split sessions), physical violations (independent energy/duration models produce impossible states like 50 kWh in 10 min o...

📖 Read original article


230. CFR without Unbiasedness: Deterministic Guarantees for Persistent Public-Chance Schedules ​

Author: Jiaxing Guo, Lei Ye
Published: 8/18/2026, 4:00:00 AM
Categories: cs.GT, cs.LG

arXiv:2608.14761v1 Announce Type: cross Abstract: At a finite public-chance cut, counterfactual regret minimization (CFR) must choose how many outcomes to evaluate before each regret update. Exact evaluation processes the full cut at one strategy profile; persistent partial evaluation processes a fi...

📖 Read original article


231. Cross-Modal Ultrasound-MRI Learning for Fetal Brain Ventricular Volumetry and Abnormality Screening ​

Author: Yuhao Huang, Yuanji Zhang, Yuhuan Lu, Dong Ni, P. Ellen Grant, Davood Karimi
Published: 8/18/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG

arXiv:2608.14763v1 Announce Type: cross Abstract: Assessment of ventriculomegaly (VM) on fetal brain ultrasound relies primarily on measuring lateral ventricular atrial width on standard planes, which is operator-dependent and may not fully reflect the overall ventricular enlargement. Fetal brain MR...

📖 Read original article


232. Beyond Boundary Noise: Aggregated Aleatoric Uncertainty Fails to Capture Presence Ambiguity in 3D Lung Nodule Segmentation ​

Author: Simon Baur, Arne Schernich, Ekin B"oke, Wojciech Samek, Jackie Ma
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.14766v1 Announce Type: cross Abstract: Uncertainty estimation is critical for the safe clinical deployment of deep learning in medical image segmentation, with aleatoric uncertainty theoretically designed to capture irreducible data ambiguity. However, whether entropy-based measures refle...

📖 Read original article


233. Uncertainty Identifies Difficult Samples Across Methods: A Multi-Task Study on a Heterogeneous Skin Lesion Dataset ​

Author: Leon Koole, Jiapan Guo, Matias Valdenegro-Toro
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.14768v1 Announce Type: cross Abstract: Skin lesion classifiers can be confidently wrong on the cases that matter most, so knowing when a prediction should not be trusted is clinically as useful as the prediction. We study uncertainty quantification on a dataset pooled from many ISIC sourc...

📖 Read original article


234. From Errors to Proofs: Minimal-Core-Guided Repair for Neuro-Symbolic Constraint Solving ​

Author: Dipankar Sarkar
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.LO, cs.PL, cs.SC, math.OC

arXiv:2608.14771v1 Announce Type: cross Abstract: Making language models solve constraint problems reliably often means having them translate the problem into a formal specification and delegating the search to a sound solver. But the translation is itself a language-model task, and an unfaithful tr...

📖 Read original article


235. AMPLIFAI: A Multiphase CT Dataset for Benchmarking Clinical Reasoning in LI-RADS Assessment of Liver Lesions ​

Author: Pranav Kulkarni, Nikhil Shah, Amritansh Suryavanshi, Jana Delfino, James Tonascia, Jade Wong-You-Cheong, Barton Lane, Joseph Chirico, Jeffrey D. Hirsch, Ang Li, Heng Huang, Florence X. Doo
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.14778v1 Announce Type: cross Abstract: Hepatocellular carcinoma (HCC) is the third leading cause of cancer-related mortality worldwide, with early detection improving survival from 70%. The standardized LI-RADS criteria establish a biopsy-free, fully imaging-based framework that can serv...

📖 Read original article


236. Prompting is not enough: supervised baselines and leakage control for measuring shared decision-making with LLMs in pediatric encounters ​

Author: Bernardo Modenesi, Jody Lin, Kimberly Kaphingst, Angela Zhu, Maya Wheeler, Peilu Zhang, Angela Fagerlin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.14792v1 Announce Type: cross Abstract: Objectives: To determine whether zero-shot prompting of a large language model (LLM) is sufficient to detect shared decision-making (SDM) behaviors in real clinical encounters, and whether supervised learning adds value under patient-grouped, nested ...

📖 Read original article


237. Zero-Shot Adaptation of Medical Vision Foundation Models for High-Frequency Micro-Ultrasound Prostate Segmentation ​

Author: Ayusha Abbas, Saram Abbas, Kabita Adhikari
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.14796v1 Announce Type: cross Abstract: Prostate cancer claims a life every 80 seconds. Early detection is needed to prevent disease progression, and both PSA density calculation and biopsy decisions rely on knowing the exact boundary of the gland. Conventional ultrasound at 6-12 MHz blurs...

📖 Read original article


238. Do LLMs Know What to Ask and When? Evaluating Multi-Turn Information Seeking ​

Author: Yepeng Huang, Jiawen Zhang, Michelle Dai, Xiaorui Su, Shanghua Gao, Zi Wang, Marinka Zitnik
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.14808v1 Announce Type: cross Abstract: When a user question is underspecified, a capable model should recognize that its context is insufficient, identify the missing information, ask for it, and respond only once that information determines a unique answer. We formalize multi-turn inform...

📖 Read original article


239. What Makes a Good Layer? Assessing the Layer-Wise Intrinsic Properties of Music Foundation Models ​

Author: Angelos-Nikolaos Kanatas, Yuexuan Kong, Pablo Alonso-Jim'enez, Xavier Serra, Dmitry Bogdanov
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS

arXiv:2608.14819v1 Announce Type: cross Abstract: Music foundation models are commonly used as frozen audio feature extractors, yet selecting which layer to extract from remains largely heuristic. Current practice defaults to fixed depths or multi-layer fusion, with limited understanding of why cert...

📖 Read original article


240. A Parameter-Free Few-Shot Evaluation for Elephant Vocalisation Classification ​

Author: Christiaan M. Geldenhuys, Thomas R. Niesler
Published: 8/18/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, cs.SD, q-bio.QM

arXiv:2608.14824v1 Announce Type: cross Abstract: We present a parameter-free episodic evaluation of nearest-centroid classification for elephant vocalisations on fixed pretrained acoustic embeddings, across the Elephant Voices (EV) and Linguistic Data Consortium (LDC) datasets. Rather than asking w...

📖 Read original article


241. Explainability Boosted Anomaly Detection Framework for O-RAN based NextG Networks ​

Author: Nurullah Aksu, Ali Fuat Sahin, Semiha Tedik Ba\c{s}aran
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.CR, cs.LG

arXiv:2608.14826v1 Announce Type: cross Abstract: The wireless networks have historically faced significant security vulnerabilities, necessitating advanced anomaly detection mechanisms, especially as networks evolve towards 6G and beyond. This study introduces an advanced anomaly detection framewor...

📖 Read original article


242. MINT: Min-Selection Preference Distillation for Balanced Multi-Objective Alignment ​

Author: Tony Tu, Sayan Chakraborty, Ruomeng Xu, Tony Qin, Austin Tian
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.14828v1 Announce Type: cross Abstract: Aligning a language agent to several objectives at once is a persistent failure mode of preference-based training: when objectives are combined additively, optimization collapses onto whichever is cheapest to improve and sacrifices the rest, so a sup...

📖 Read original article


243. Discovering High-Quality Chess Puzzles with Offline Reinforcement Learning ​

Author: Allen Nie, Anirudhan Badrinath, Nicholas Tomlin, Timothy Dai, Carissa Yip, Rose E Wang, Emma Brunskill, Chris Piech
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.14851v1 Announce Type: cross Abstract: Learning and skill mastery require extensive and deliberate practice. In many learning settings, producing high-quality pedagogical materials can require a high level of domain expertise and be very time-consuming. Pedagogical materials often need to...

📖 Read original article


244. BRAID: Learning Equilibrium Maps in Interdependent Security Games via Weight-Tied Iterative Graph Neural Networks ​

Author: Elnaz Nowrouzi, Zhiqun Zuo, Xueru Zhang, Mohammad Mahdi Khalili
Published: 8/18/2026, 4:00:00 AM
Categories: cs.GT, cs.LG

arXiv:2608.14856v1 Announce Type: cross Abstract: Computing Nash equilibria in interdependent security (IDS) games on networks is computationally expensive: best-response dynamics may need hundreds of iterations per instance, and downstream tasks such as auditing, stress-testing, and incentive desig...

📖 Read original article


245. STAR-FL: Secure Federated Learning with Spatial-Temporal Analysis and Robust Aggregation ​

Author: Nawrin Tabassum, Yanzhao Wu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.14861v1 Announce Type: cross Abstract: Data poisoning attacks pose serious security threats to Federated Learning (FL) systems in Computer Vision. Despite growing research attention, two key challenges remain for existing defense techniques: (1) accurately distinguishing between benign an...

📖 Read original article


246. ARISE: An adaptive residual-informed stability ensemble for feature selection in small-sample biomedical omics ​

Author: Zardad Khan, Amjad Ali, Naz Gul, Sheema Gul, Saeed Aldahmani
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.14866v1 Announce Type: cross Abstract: Objective: Small-sample molecular classification requires feature selectors that identify predictive, stable, and nonredundant subsets for binary and multiclass outcomes. We propose ARISE (Adaptive Residual-Informed Stability Ensemble), which integra...

📖 Read original article


247. When do machine-learned exchange-correlation improvements inherit into density-functional tight binding? ​

Author: Can Polat, Mustafa Kurban, Erchin Serpedin, Hasan Kurban
Published: 8/18/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG, physics.chem-ph, quant-ph

arXiv:2608.14875v1 Announce Type: cross Abstract: Machine-learned exchange-correlation functionals correct band gaps at near-semilocal cost, while density-functional tight binding reaches the $10^3$-$10^6$-atom regime; combining them assumes that a better parent yields a better parameterization, but...

📖 Read original article


248. Workspace Topology as an Attack Vector in Agentic Coding Assistants ​

Author: Alexandre G. R. Day, Pradeep Yadlapalli, Sriram Venkatapathy, Thomas Paniagua, Nick Raines, Sahil Wadhwa, Himanshu Kumar, Andy Luo, Sudeep Panyam, Rikhiya Ghosh, Pranab Mohanty, Giri Iyengar
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG

arXiv:2608.14876v1 Announce Type: cross Abstract: Agentic coding assistants are finding widespread use, not just in new code development but in quickly ingesting and leveraging third-party code. This opens up a risk of malicious code being ingested as these coding tools operate with broad filesystem...

📖 Read original article


249. Interpretable Cross-Lingual Alignment in Small Language Models: Probing Cultural and Pragmatic Reasoning in Japanese-English Bilingual LLMs ​

Author: Florian Braun
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.14896v1 Announce Type: cross Abstract: Large language models work well on English and behave in poorly understood ways on languages typologically far from it. Japanese is a clean example, where evaluation still leans on translation quality and JGLUE-style benchmarks, which roll lexical, s...

📖 Read original article


250. Optimal Watermark Localization in Mixed-Source Large Language Model Texts ​

Author: Jose H. Blanchet, T. Tony Cai, Xiang Li, Hao Liu, Qi Long, Weijie J. Su
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ME, cs.CL, cs.LG, stat.ML

arXiv:2608.14906v1 Announce Type: cross Abstract: Watermarking provides a principled way to authenticate text generated by large language models (LLMs). In practice, however, the final text may be mixed-source, with watermark evidence surviving at only a subset of token positions after rewriting, in...

📖 Read original article


251. Distinguishing AI-Generated Music from Edited Audio as a Hard-Negative Robustness Task ​

Author: Alexandru-Stefan Morosanu, Valerian Cecan, Stefan-Daniel Achirei, Laura Erhan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG

arXiv:2608.14916v1 Announce Type: cross Abstract: AI-generated music detectors are commonly evaluated against original songs, but real-world uploads are often remixed, re-encoded, pitch-shifted, or otherwise edited. These edited versions form a difficult negative class: they are not generated by AI,...

📖 Read original article


252. LLMs Can Predict Failure Risk, But Struggle to Predict Which Collaboration Protocol Pays Off: Cost-Aware Protocol Routing Across Reasoning Tasks ​

Author: Chih-Hsuan Yang, Jingyan Jiang, Cheng-Hau Yang, Vikram Vasudevan, Huihuo Zheng, Venkatram Vishwanath, Rajeev Thakur
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.14927v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) systems can improve reasoning by spending more computation, but deployment requires deciding when extra collaboration is worth its cost. We isolate this decision by running every problem under four protocols whi...

📖 Read original article


253. Training Leaves Traces: Centered Residual Signatures for Language Model Lineage Verification ​

Author: Aman Singh Thakur, Rayan Khoury
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.14929v1 Announce Type: cross Abstract: Open-weight language models are fine-tuned, quantized, pruned, and merged, yet their provenance is often undocumented. We study data-free white-box lineage verification: can weights alone reveal whether two compatible model checkpoints share ancestry...

📖 Read original article


254. Developing an Offshore Machine Learning Surface Layer Scheme ​

Author: Susan Dettling, Sue Ellen Haupt, Thomas Brummet, Patrick Hawbecker, Branko Kosovi'c, David John Gagne
Published: 8/18/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG

arXiv:2608.14935v1 Announce Type: cross Abstract: Turbulent fluxes between the surface and the atmosphere are typically parameterized using empirically fit relationships. Here we test machine learning techniques for fitting the relationship for the offshore environment. To do that, data from three o...

📖 Read original article


255. SkillComposer: Learning Reusable Skills for Natural-Language Robot Programming ​

Author: John Woods, Hasti Seifi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.RO, cs.CL, cs.LG

arXiv:2608.14944v1 Announce Type: cross Abstract: Natural-language interfaces can lower the barrier to programming robots, but existing systems struggle when users request complex tasks. While large language models (LLMs) perform well with simple commands, they often struggle to generate code for mu...

📖 Read original article


256. T-LLM Compiler: Trusted LLM-based Code Optimization and Verification Framework ​

Author: Zahra Fazel, Sunanda Gamage, Shayan Shirahmad Gale Bagi, Amir H. Ashouri, Tomasz S. Czajkowski, Bryan Chan, Reza Azimi, Yaoqing Gao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.PF, cs.PL

arXiv:2608.14953v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have opened opportunities to apply high-level code transformations to the field of code optimization, and it has since emerged as one of the most fundamental tasks for LLMs to perform; however, at prese...

📖 Read original article


257. A Deep Learning Model for Spatially Clustered Data via Differentiable Cluster Assignment ​

Author: Kexuan Li, Weidong Ma
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.14968v1 Announce Type: cross Abstract: We consider nonparametric regression when the association between a response and its covariates changes across an unknown partition of a spatial domain. The proposed estimator learns the partition and the cluster-specific regression functions jointly...

📖 Read original article


258. Spinning Conformal Correlators from Neural Networks ​

Author: Manas Dogra, James Halverson, Joydeep Naskar
Published: 8/18/2026, 4:00:00 AM
Categories: hep-th, cs.LG

arXiv:2608.15001v1 Announce Type: cross Abstract: We construct spinning conformal fields from neural networks and the embedding formalism, computing their two-, three- and four-point functions in examples, building on scalar conformal field techniques introduced in \cite{Halverson:2024axc}. For a pa...

📖 Read original article


259. NPU Offloading of a Frozen Visual Encoder for Robot Policy Training ​

Author: Hyojun Yun, Seungjae Won, Hyungpil Moon
Published: 8/18/2026, 4:00:00 AM
Categories: cs.RO, cs.AR, cs.LG

arXiv:2608.15002v1 Announce Type: cross Abstract: When a robot policy is trained for a new task or dataset, its visual encoder can be frozen and only its action generation module trained, reducing training cost. Freezing removes the encoder's backward pass, but its forward pass must still run at eve...

📖 Read original article


260. DualMiT-Net: Local-Global Transformer-Convolutional Fusion for Breast Mass Segmentation in Mammographic Regions of Interest ​

Author: Alibek Kamiluly, Milana Muratova, Yash Patel, Fan Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.15019v1 Announce Type: cross Abstract: Breast mass segmentation is an important step in computer-aided mammography, but it remains difficult because masses can have low contrast, irregular shapes, and boundaries that blend with surrounding breast tissue. To address this problem, we presen...

📖 Read original article


261. Prototype-Rectified Iterative Self-supervised Manifold Denoising under Severe Acoustic Shift ​

Author: Ashish Anand Shukla, Rini Smita Thakur, Aryan Das, Vinod K. Kurmi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SD, cs.LG

arXiv:2608.15037v1 Announce Type: cross Abstract: Audio-Text Foundation Models (ATMs) fail catastrophically under severe acoustic noise, yet existing adaptation strategies either rely on gradient-based Test-Time Adaptation (TTA), which reinforces noise rather than signal, or on prompt tuning that re...

📖 Read original article


262. Uncovering Hidden Leptonic Correlations with Flow Matching and Autoencoders ​

Author: Haruto Kitagawa, Satsuki Nishimura, Hajime Otsuka
Published: 8/18/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-th

arXiv:2608.15042v1 Announce Type: cross Abstract: We perform a global search for values of the Yukawa matrices and Majorana masses in the Type-I seesaw mechanism. Using flow matching, which is a generative artificial intelligence (generative AI) method, we generate a broad set of solutions reproduci...

📖 Read original article


263. RecurrentGPT: Expressive Depth through Recurrent Modulation in Transformers ​

Author: Amr Hegazy, Amr Alanwar, Mostafa Elhoushi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.15062v1 Announce Type: cross Abstract: Scaling transformer language models creates an inherent tension between expressivity and memory efficiency. While unique weights across layers preserve functional specialization---from input-grounding to abstract refinement---they incur a substantial...

📖 Read original article


264. Distribution-free false-alarm calibration and chance-corrected spatial evaluation for industrial anomaly detection ​

Author: Jie Deng
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.15090v1 Announce Type: cross Abstract: Studies of industrial visual inspection commonly report the area under the receiver operating characteristic curve (AUROC) and the overlap between anomaly maps and defect masks. Neither measure specifies the false-alarm rate at a selected threshold, ...

📖 Read original article


265. Fast Test-Time Refinement for Robust Learned Image Compression ​

Author: Jiaming Liang, Chi-Man Pun, Weisi Lin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.15113v1 Announce Type: cross Abstract: Learned image compression (LIC) has demonstrated remarkable rate-distortion (RD) performance in benign settings. However, the high representational capacity endowed by deep neural networks (DNNs) comes at the expense of increased adversarial vulnerab...

📖 Read original article


266. Anatomy of a Quantized Agent: VRAM Stability and Forecasting in Code-Synthesis Agentic Workloads ​

Author: Anubhab Banerjee
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.DC, cs.LG

arXiv:2608.15117v1 Announce Type: cross Abstract: Analytical models of peak VRAM consumption for LLM inference decompose memory into weight-storage, KV-cache, and activation terms parameterized by step count, tool invocations, and context expansion. We evaluate this decomposition empirically within ...

📖 Read original article


267. Sufficient Dimesion Reduction via Generalized Stein's Lemma ​

Author: Ye Tian
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.15121v1 Announce Type: cross Abstract: Sufficient dimension reduction (SDR) seeks the minimal subspace of the predictors that captures the full conditional distribution of the response, which is known as the central subspace (CS). When the response is multivariate, the problem becomes con...

📖 Read original article


268. Scale-Consistent Posterior Dynamics for Diffusion Inverse Problems ​

Author: Zhaoqiang Liu, Tongyao Pang, Ruibing Wang, Yang Zheng
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2608.15144v1 Announce Type: cross Abstract: Posterior sampling with a pretrained diffusion prior is governed by a conditional score whose intermediate likelihood component is generally intractable. We begin from an ideal one-parameter posterior SDE family in which a stochasticity parameter con...

📖 Read original article


269. An Adaptive Gradient Clipping and Noise Injection Mechanism for Differentially Private Federated Learning ​

Author: Wenjing Wei, Alla Jammine, Farid Nait-Abdesselam
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.15153v1 Announce Type: cross Abstract: Differentially private federated learning must balance privacy protection against model accuracy and training efficiency. Static gradient clipping applies a fixed threshold throughout training and across model layers, which can cause excessive clippi...

📖 Read original article


270. Beyond Effective Sample Size: Effective Number of Proposals for Adaptive Importance Sampling ​

Author: Ali Mousavi, Victor Elvira
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.15154v1 Announce Type: cross Abstract: Population-based adaptive importance sampling (AIS) methods use a set of proposal densities to approximate complex target distributions. Their performance is commonly assessed through effective sample size (ESS) and related weight-based diagnostics, ...

📖 Read original article


271. Identifying parameter couplings and uncertainties of mixed-noise stochastic systems via full-covariance Gaussian mixture network ​

Author: Xiaolong Wang, Xiangwen Hao, Jing Feng, Yuanyuan Liu, Yong Xu
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, physics.comp-ph

arXiv:2608.15198v1 Announce Type: cross Abstract: Parameter identification of stochastic dynamical systems driven by mixed noises is challenging due to intractable likelihood functions. We propose PENN-GMD, a parameter estimation neural network that maps partially observed trajectories to a Gaussian...

📖 Read original article


272. The Distributional View of Knowledge Distillation ​

Author: Gordei Verbii, Juho Lee
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.15215v1 Announce Type: cross Abstract: Token-level knowledge distillation (KD) matches two conditional distributions per position, yet the standard objectives compare them pointwise: a Kullback-Leibler gradient is blind to which wrong token receives probability mass. We develop a distribu...

📖 Read original article


273. Decentralized Federated Learning for Heterogeneous Multi-Task Semantic Communication ​

Author: Lin Yin, Tiejun Lv, Weicai Li, Xi Yu, Xiaoyu He
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.15256v1 Announce Type: cross Abstract: Collaborative training in distributed semantic communication (DSC) networks typically relies on decentralized federated learning (DFL). However, pushing topology-agnostic aggregation into heterogeneous, multi-task environments creates a fundamental b...

📖 Read original article


274. BrainLinear: A Linear Model for Brain Network Analysis in Sparse Tangent Subspaces ​

Author: Sijing Wu, Dongyuan Li, Miaoting Huang, Weiwei Ye, Ying Zhang, Feng Xia, Renhe Jiang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.GR, cs.LG

arXiv:2608.15266v1 Announce Type: cross Abstract: Functional connectome analysis examines brain-region interactions to understand and identify disorders such as autism spectrum disorder and Alzheimer's disease. Existing methods typically use GNNs and Transformers to model the full functional connect...

📖 Read original article


275. Memory-Bounded Continuation of Greedy Sampling for Continual Anomaly Detection ​

Author: Yoon Gyo Jung, Jaewoo Park, Kuan-Chuan Peng, Seongdeok Bang, Octavia Camps
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.15277v1 Announce Type: cross Abstract: Greedy sampling produces a compact yet representative summary of normal data, which is essential for reliable anomaly detection that relies on measuring distance from normality. For continual anomaly detection where tasks arrive sequentially, extendi...

📖 Read original article


276. Convolution Smoothed Quantile Regression for XGBoost ​

Author: Mandy Yao (University of Toronto), Meredith Franklin (University of Toronto)
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.15290v1 Announce Type: cross Abstract: The increasing availability of large and complex datasets across many scientific disciplines has led to widespread adoption of machine learning (ML) for prediction. However, most ML algorithms focus on point estimation and provide limited information...

📖 Read original article


277. A Unified Geometric Framework for Developmental Analysis of Spatial Transcriptomic Data ​

Author: Mary Chriselda Antony Oliver, Kaitlyn Hohmeier, Tuyen Tran, Alejandra Castillo, Caroline Moosm"uller, Shiying Li
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.MG

arXiv:2608.15306v1 Announce Type: cross Abstract: High-throughput single-cell and spatial transcriptomic technologies provide high-resolution snapshots of heterogeneous cellular states, but their destructive nature prevents repeated measurements of the same cells over time. Consequently, temporal an...

📖 Read original article


278. FedADB: Class Anchor-Driven Dual-Branch Federated Learning for Mitigating Forgetting ​

Author: Zhenyan Liu, Hua Zhang, Haoran Gao, Qi Li, Hongliang Zhu, Huiyu Zhou, Zongliang Shen, Yanxin Xu, Jiahui Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.15310v1 Announce Type: cross Abstract: Multimodal data collected by heterogeneous devices are used for collaborative training, where federated learning (FL) serves as a key paradigm for effective distributed modeling with data privacy preservation. However, local training suffers from the...

📖 Read original article


279. The Physical Cutoff Does Not Restore Homogenization: Phase-Dependent Burning in the Strain G-Equation ​

Author: Michele Caprio
Published: 8/18/2026, 4:00:00 AM
Categories: math.AP, cs.LG

arXiv:2608.15337v1 Announce Type: cross Abstract: We disprove the expectation stated by Xin, Yu, and Ronney that the physical positive part strain $G$-equation should possess an effective burning velocity in cellular flows. For the standard cellular flow in dimension two $V_A(x_1,x_2)=A(-\sin x_1\co...

📖 Read original article


280. Prediction Inference of Time Series with Standard ReLU Deep Neural Networks ​

Author: Kejin Wu
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO

arXiv:2608.15362v1 Announce Type: cross Abstract: We propose a methodology based on the standard ReLU Deep Neural Networks (DNN) to make predictions and quantify their uncertainty. Classically, people rely on linear, non-linear, or non-parametric kernel methods to fit and then predict the time serie...

📖 Read original article


281. FedPA-LoRA: Product-Aligned Framework for Mitigating Aggregation and Initialization Errors in Heterogeneous Federated LoRA ​

Author: Juseok Jeon, Ramy E. Ali, Doyun Kwon, Myungbeom Her, Jinhwi Kim, Jinhyun So
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.15381v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) enables efficient federated fine-tuning of large language models, but its factorized parameterization creates a tension between accurate aggregation of local updates and continuity of locally optimized factors. Factor-wise ...

📖 Read original article


282. ETHOS: Towards a Modular Ethics Framework for Clinical Multi-Agent Systems ​

Author: Rakesh Sharma, Sydney Pugh, Cameron Beeche, Pankhuri Singhal, Rachel Wu, Margaret Eby, Jeffrey Duda, James Gee, Kyra O'Brien, Hersh Sagreiya, Marina Serper, Victoria Gershuni, Angela Bradbury, Anurag Verma, Eric Eaton, Kevin B. Johnson, Walter Witschey
Published: 8/18/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.LG

arXiv:2608.15424v1 Announce Type: cross Abstract: The rapid adoption of large language models has enabled the development of clinical multi-agent systems (MAS) capable of integrating multimodal patient data and supporting increasingly complex clinical decision-making. However, the deployment of thes...

📖 Read original article


283. NeuRoute: Logit-Guided Neural Routing for Billion-Scale Vector Search with Sub-Hour Index Construction ​

Author: Xingqiao Wang, Zi Wang, Xiaowei Xu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DB, cs.IR, cs.LG

arXiv:2608.15438v1 Announce Type: cross Abstract: Building approximate nearest neighbor (ANN) indexes at billion scale is often dominated by expensive global clustering or graph construction, making time-to-index a first-order systems concern. We present NeuRoute, a learned hashing index that turns ...

📖 Read original article


284. Language models suffer from a curse of ambiguity ​

Author: Nicolas Zucchet, Hyun Dong Lee, Scott Linderman
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.NE

arXiv:2608.15448v1 Announce Type: cross Abstract: Large language models increasingly rely on sampling as a driver of their own improvement, making the fidelity of their learned distributions more critical than ever. Yet, not all distributions are equally easy to learn. In this work, we identify a cu...

📖 Read original article


285. Maintaining IoT Device Identification under Concept Drift via Budget-Aware Traffic Labeling ​

Author: Shayan Azizi, Norihiro Okui, Masataka Nakahara, Ayumu Kubota, Gustavo Batista, Hassan Habibi Gharakaheili
Published: 8/18/2026, 4:00:00 AM
Categories: cs.NI, cs.CR, cs.LG

arXiv:2608.15465v1 Announce Type: cross Abstract: Identification of IoT device types from passive traffic is increasingly used for security management in enterprise and ISP networks. However, the performance of machine learning-based classifiers gradually degrades under concept drift as device behav...

📖 Read original article


286. Do Language Models Consistently Encode the Current Year? ​

Author: Suze van Adrichem, Aditi Bhaskar, Diyi Yang, Christopher Potts, Jing Huang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.15507v1 Announce Type: cross Abstract: A consistent concept of the current time is important for temporal reasoning, yet how language models represent the current time is not well understood. We contribute two tasks that probe the current year in conceptually distinct ways: an associative...

📖 Read original article


287. Temporal Logic Guided Universal Task Representations for Reinforcement Learning ​

Author: Hao Zhang, Zhangli Zhou, Zhen Kan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.RO, cs.FL, cs.LG

arXiv:2608.15509v1 Announce Type: cross Abstract: Task guided agents demonstrate strong performance in a wide range of complex tasks. However, most existing task representation algorithms are tailored to specific contexts and struggle to generalize across diverse scenarios. Moreover, they typically ...

📖 Read original article


288. DeltaLog: Deferred Materialization of Recurrent States for Linear Attention Decoding ​

Author: Junqing Lin, Jingwei Sun, Guangzhong Sun
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.15533v1 Announce Type: cross Abstract: Linear attention models eliminate the quadratic prefix computation and context-growing KV cache of softmax attention by replacing pairwise token interactions with recurrent state updates. However, existing decoding implementations often materialize a...

📖 Read original article


289. L3Cube-IndicQuest v2: A Large-Scale Multilingual Benchmark for Evaluating Factual Knowledge of Large Language Models Across Indic Languages ​

Author: Rinit Jain, Tirthraj Mahajan, Advait Joshi, Raviraj Joshi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.15535v1 Announce Type: cross Abstract: We present L3Cube-IndicQuest v2, a large-scale gold-standard multilingual question-answering benchmark for evaluating the India-specific factual knowledge of Large Language Models (LLMs). The benchmark comprises 3,471 curriculum-grounded English ques...

📖 Read original article


290. RigidBench: Evaluating Rigid-Body Physics in Video Generation Models ​

Author: Swarnim Jain, Shangzhe Wu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.15555v1 Announce Type: cross Abstract: Video models are increasingly used to predict what happens next in a scene, yet the metrics commonly used to compare their outputs say little about whether the predicted objects move correctly. Motion, geometry, identity, background stability, and vi...

📖 Read original article


291. A Counterexample to the Tang Zhang Schatten Norm Conjecture and Sharp Positive Results ​

Author: Zijian Zeng, Houde Liu, Kurunathan Ratnavelu
Published: 8/18/2026, 4:00:00 AM
Categories: math.CO, cs.LG, math.FA

arXiv:2608.15558v1 Announce Type: cross Abstract: For $m\geq 2$, let $c_p(m)$ be the all-dimensional best constant in $$ \left|\sum_{k=1}^m A_k\right|p \leq c_p(m)\left|\sum^m |A_k|\right|_p. $$ Tang and Zhang conjectured an explicit formula for every finite $p>1$. We disprove the conject...

📖 Read original article


292. Rotation-Invariant Multi-IMU Activity Recognition under Independent Per-Location Orientation Shifts ​

Author: Seungyeol Baek, Yoonbyung Chai, Yonghyeon Lee, Sungjoon Choi, Sungho Suh
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.15621v1 Announce Type: cross Abstract: Human Activity Recognition (HAR) with self-administered wearables, such as at-home rehabilitation and exercise monitoring, often requires reattaching inertial measurement units (IMUs) across sessions. In multi-IMU settings, this can induce independen...

📖 Read original article


293. When Is Shallow Enough? Adaptive Split Federated Learning with Client-Specific Sufficiency Estimation ​

Author: Wenhao Yuan, Chenchen Lin, Wenhao Hu, Jian Chen, Jinfeng Xu, Shujie Li, Edith Cheuk Han Ngai
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG

arXiv:2608.15639v1 Announce Type: cross Abstract: \textit{Split Federated Learning} (SFL) enables distributed model training by splitting networks between the server and clients. However, under client heterogeneity, the conventional static split strategy may be suboptimal because clients can differ ...

📖 Read original article


294. On Stopping Rules and Spatial Adaptation for CART ​

Author: Zineng Xu, Yuchao Cai, Yan Shuo Tan
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2608.15649v1 Announce Type: cross Abstract: The popular CART algorithm for regression trees combines a greedy splitting rule with a stopping rule, but while the splitting rule has been well studied, the statistical role of stopping rules is less well understood. Meanwhile, although regression ...

📖 Read original article


295. Adaptive Heterogeneous Compression for Resource-Efficient Federated Knowledge Distillation ​

Author: Chenwang Liu, Yijun Liu, Chang Liu, Xu Zhang, Pengchao Han
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.15660v1 Announce Type: cross Abstract: Federated learning (FL) enables privacy-preserving distributed model training but faces challenges from heterogeneous model architectures and limited communication resources at the network edge. Federated knowledge distillation (FedKD) alleviates mod...

📖 Read original article


296. Integrating Persuasion Theory into the Epidemiological Modelling of Health Misinformation Spread on Social Media ​

Author: Mkululi Sikosana, Sean Maudsley-Barton, Oluwaseun Ajao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SI, cs.AI, cs.CL, cs.LG

arXiv:2608.15689v1 Announce Type: cross Abstract: This study presents a hybrid epidemiological and behavioural framework to simulate the spread of health misinformation on social media. We extend the classical Susceptible--Infected--Recovered (SIR) model to a six-compartment structure (SIRMMM), inco...

📖 Read original article


297. Adding Voice Cloning to Text-to-Audio-Video Models with a Single Zero-Initialised Layer ​

Author: Ivan Mikheev, Viacheslav Vasilev, Anna Dmitrienko, Alexey Letunovskiy, Ivan Kirillov, Kirill Chernyshev, Denis Dimitrov
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG, cs.MM

arXiv:2608.15690v1 Announce Type: cross Abstract: Text-to-audio-video (T2AV) generation models produce a video and its soundtrack from a textual description, but offer no control over whose voice speaks in the output. We show that a base T2AV model can be turned into a voice-cloning model by adding ...

📖 Read original article


298. BERTopic-Virality Prioritisation: A Scalable Framework for Thematic and Comparative Analysis of COVID-19 and Monkeypox Misinformation on Twitter ​

Author: Mkululi Sikosana, Sean Maudsley-Barton, Oluwaseun Ajao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SI

arXiv:2608.15691v1 Announce Type: cross Abstract: Health misinformation circulating during pandemics can gain traction rapidly, creating harmful narratives that compete with public health guidance. Most topic-modelling pipelines treat engagement as an external outcome, limiting their ability to prio...

📖 Read original article


299. Large Models for Small Devices: Recent Advances and Empirical Analysis of Edge AI Deployment ​

Author: Subhransu Das, Jiaming Cheng, Arnav Kumar, Sadia Afrose, Mingzhe Han, Michael Silagy, Shreya Palande, Brijesh Soni, Rajiv Ramnath
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.15693v1 Announce Type: cross Abstract: Running large AI models on resource-constrained edge devices requires model compression to reduce model size and computation. What compresses well, however, need not deploy well. We survey dozens of recent works that report compression results on rea...

📖 Read original article


300. Deep learning-based computed tomography (CT) derived body composition classifier for colorectal cancer patients ​

Author: Eve Harling (James Watt School of Engineering, College of Science & Engineering, University of Glasgow, Glasgow, UK), Chattarin Pumtako (Academic Unit of Surgery, School of Medicine, College of Medical Veterinary & Life Sciences, University of Glasgow, Glasgow, UK), Bernd Porr (James Watt School of Engineering, College of Science & Engineering, University of Glasgow, Glasgow, UK), Donald C McMillan (Academic Unit of Surgery, School of Medicine, College of Medical Veterinary & Life Sciences, University of Glasgow, Glasgow, UK), Ross D Dolan (Academic Unit of Surgery, School of Medicine, College of Medical Veterinary & Life Sciences, University of Glasgow, Glasgow, UK)
Published: 8/18/2026, 4:00:00 AM
Categories: eess.IV, cs.LG

arXiv:2608.15712v1 Announce Type: cross Abstract: Background: Accurate body composition analysis using Computed Tomography (CT) scans is essential for assessing skeletal muscle area (SMA) and skeletal muscle density (SMD), key markers of nutritional status in cancer patients. Conventional manual met...

📖 Read original article


301. Continuous Quantum Feedback Control via Kraus-Parameterized Belief Reinforcement Learning ​

Author: Priyanshi Singh, Krishna Bhatia
Published: 8/18/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.15715v1 Announce Type: cross Abstract: Quantum feedback control requires acting on noisy continuous measurement records without direct access to the underlying quantum state. We propose Kraus-Parameterized Belief Reinforcement Learning, a pipeline in which a recurrent encoder, constrained...

📖 Read original article


302. PLeDO: Pain Level Detection for Osteoarthritis from EMR Data ​

Author: Yuhao Chen, Jiahao Cai, Nafiz Sadman, Farhana Zulkernine, John Queenan, David Barber
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.ET, cs.IR, cs.LG

arXiv:2608.15719v1 Announce Type: cross Abstract: Osteoarthritis (OA) is a progressive chronic joint disease resulting in a breakdown of articular cartilage and bone when damaged joint tissues are not able to normally repair themselves. The aim of this pilot research study is to understand the pain ...

📖 Read original article


303. Machine Learning Approaches to Decoding Topological Quantum Codes ​

Author: Changwon Lee, Tak Hur, Jeongwoo Jae, Daniel K. Park
Published: 8/18/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.15760v1 Announce Type: cross Abstract: Decoding is an essential component of quantum error correction (QEC), translating stabilizer measurement outcomes into corrective actions that suppress logical errors and preserve logical quantum information. Building fault-tolerant architectures req...

📖 Read original article


304. Provenance, Not Behaviour: A Serialisation Artifact in Edge-IIoTset and a Leakage-Free Benchmark for Precision-Agriculture Intrusion Detection ​

Author: Mostafa M. Galal
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.15761v1 Announce Type: cross Abstract: Edge-IIoTset is the reference benchmark for machine-learning intrusion detection in the industrial Internet of Things, and results reported on it cluster above 99%. We show that much of that performance is not intrusion detection. The preprocessing r...

📖 Read original article


305. Global Simulation-Guided Dynamic Operator Scheduling for Efficient Multi-Tenant Model Serving ​

Author: Weinan Liu, Zeyuan Ding, Dian Ding, Chengcheng Wan, Lu Tang, Guangtao Xue, Jiwu Shu, Yiming Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.OS, cs.DC, cs.LG

arXiv:2608.15762v1 Announce Type: cross Abstract: Container-granularity scheduling leaves abundant short-lived idle slices within containers unexploited. Reallocating containers is too heavyweight to utilize such fine-grained opportunities under SLA constraints, and operator-level scheduling require...

📖 Read original article


306. Inferential Evaluation of Surrogate-Derived Models under Covariate Shift ​

Author: Longtian Shi, Molei Liu, Doudou Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP, stat.ME

arXiv:2608.15783v1 Announce Type: cross Abstract: In transfer-learning settings, a model derived from abundant surrogate labels may be deployed in a target population where gold-standard outcomes are unobserved. Evaluating its target performance is essential for determining whether decisions based o...

📖 Read original article


307. PWLR: Pairwise Witness Local Rejection for Boundary-Aware Out-of-Distribution Detection ​

Author: Chengyao Jia, Ruixuan Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.15802v1 Announce Type: cross Abstract: Out-of-distribution (OOD) detection remains challenging for image classifiers, especially when near-OOD samples lie close to in-distribution (ID) class boundaries. Recent vision-language detectors improve OOD detection through class semantics, local ...

📖 Read original article


308. QuantumPhaseNet: A Gauge-Covariant Geometric and Quantum-Spectral Theory of Semantic Concept Hierarchies with Prototype Validation of a Classical Quantum-Inspired Model ​

Author: Kiyotaka Kasubuchi, Kazuo Fukiya
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.15820v1 Announce Type: cross Abstract: We present QuantumPhaseNet, a gauge-covariant geometric and quantum-spectral extension of Transformer representations. Context-dependent semantic states are modeled as complex amplitudes; a covariant phase rate induces a semantic wavelength used as a...

📖 Read original article


309. How Many Samples Are Needed to Determine Causal Direction? Sharp Minimax Bounds for Bivariate LiNGAM ​

Author: Jikai Jin
Published: 8/18/2026, 4:00:00 AM
Categories: math.ST, cs.LG, econ.EM, stat.ML, stat.TH

arXiv:2608.15840v1 Announce Type: cross Abstract: We study how many observations are needed to determine the causal direction between two linearly related variables. Classical LiNGAM theory shows that independent non-Gaussian disturbances identify the direction, but does not quantify the difficulty ...

📖 Read original article


310. Generalized Linear Bandits with Memory ​

Author: Heesang Ann, Hyunjun Choi, Taehyun Hwang, Younghoon Shin, Haeju Cheong, Min-hwan Oh
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.15848v1 Announce Type: cross Abstract: We study generalized linear bandits with memory, an endogenous non-stationary setting in which rewards depend on past actions through a finite memory matrix. Building on prior work for linear models (Clerici et al., 2024), we show that the previously...

📖 Read original article


311. CoupVisor: Strategy Optimization by Round and Challenge Decision Support ​

Author: Cris Huynh
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.GT, cs.LG

arXiv:2608.15868v1 Announce Type: cross Abstract: This paper presents CoupVisor, a decision-support system for the hidden-information card game Coup. It addresses two questions: what a player should do on each turn, and when a player should challenge an opponent's claim. The system is built around a...

📖 Read original article


312. Beyond Visual CoT: Internalized Visual Thinking for Proactive Video Reasoning ​

Author: Xiaoyu Zhu, Xinke Deng, Suresh Taddewadikar, Arnab Kumar Mondal, Zhongyu Jiang, Ian Fasel, Joerg Liebelt
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG, cs.MM

arXiv:2608.15869v1 Announce Type: cross Abstract: Multimodal large language models increasingly use visual chain-of-thought (Visual CoT) to reason about spatial, temporal, and embodied environments. By generating intermediate reasoning images, Visual CoT provides an intuitive mechanism for visual fo...

📖 Read original article


313. Large Language Models as Implicit Sociological Models: Reconstructing Voting Behaviour from Sociodemographic Profiles ​

Author: Roman Neruda, Martin Bako\v{s}, Josef \v{S}lerka, V'it Tu\v{c}ek, Petra Vidnerov'a, Gabriela Kadlecov'a
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CY, cs.CL, cs.LG

arXiv:2608.15871v1 Announce Type: cross Abstract: Large language models (LLMs) trained on large-scale internet corpora encode extensive statistical regularities about social identities, attitudes, and political behaviour. This paper introduces and evaluates a methodological framework that leverages ...

📖 Read original article


314. Resource-Efficient QUBO Formulation for Anchored Currency Arbitrage ​

Author: Eric A. F. Reinhardt, Adam J. Hauser
Published: 8/18/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.15889v1 Announce Type: cross Abstract: Currency arbitrage (CA) involves trading currencies in cycles to exploit discrepancies in market valuations. Quadratic unconstrained binary optimization (QUBO) involves minimizing a quadratic cost (energy) function of binary variables. Previous works...

📖 Read original article


315. Crystal-structure design by agentic AI in a language of motifs ​

Author: Dinh-Khiet Le, Minh-Quyet Ha, Hong-Phuc Vu-Dinh, Takashi Miyake, Hiori Kino, Hieu-Chi Dam
Published: 8/18/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG

arXiv:2608.15900v1 Announce Type: cross Abstract: Data-driven materials discovery interpolates more reliably than it extrapolates and seldom reaches new structure types. We present MatEvolve, an agentic-AI framework designing crystals, proposing each candidate with a stated rationale and testing it....

📖 Read original article


316. $S^3$: A Smooth Simulation Surrogate for Optimizing Discrete Abstractions of Dynamical Systems ​

Author: Jordan Peper, James Mathias Gast, Vignesh Nanduri, Tanmayee Maram, Ethan Howes, Ivan Ruchkin
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2608.15920v1 Announce Type: cross Abstract: Intelligent systems are increasingly deployed in safety-critical settings with black-box controllers, including neural networks. The properties and behaviors of these end-to-end systems can be studied with abstraction-based methods that replace them ...

📖 Read original article


317. The Null Token Knows: Reducing Message-Free Hallucination in ASR and NMT ​

Author: Kirill Borodin, Vasiliy Kudryavtsev, Ivan Viakhirev
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SD

arXiv:2608.15940v1 Announce Type: cross Abstract: Modern encoder-decoder systems can produce fluent text even when their input contains no recoverable message. We study this failure in ASR and NMT through the models' reserved null tokens, asking whether the score for ending generation already carrie...

📖 Read original article


318. Ask to Be Sure: Informative Interactions for Confident Multi-Turn LLM Recommendation ​

Author: Cedar Site Bai, Duanshun Li, Zhenyu Liao, Sheikh Sarwar, Huiyuan Chen, Yuan Chen, Changhe Yuan, Haiyang Zhang, Qilin Qi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CL, cs.LG

arXiv:2608.15949v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled their use as conversational recommender systems (CRS), demonstrating strong recommendation accuracy and natural dialogue. However, guiding multi-turn interactions to elicit user preferences...

📖 Read original article


319. Functional anatomy of Pythia-Herwig differences with Kolmogorov-Arnold networks ​

Author: Arghya Chattopadhyay
Published: 8/18/2026, 4:00:00 AM
Categories: hep-ph, cs.LG

arXiv:2608.15952v1 Announce Type: cross Abstract: Differences between high-energy event generators can arise at several stages of the collision simulation, from the hard scattering through parton showering and hadronization to the final event. These differences are usually summarized using observabl...

📖 Read original article


320. Solvable Sokoban Without a Solver via Diffusion ​

Author: Sina Baghal
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.GT, cs.LG

arXiv:2608.15958v1 Announce Type: cross Abstract: Deciding whether a Sokoban puzzle is solvable is PSPACE-complete (Culberson, 1997): solutions can be exponentially long and there is no short certificate to check. Solvability is also a fragile property, since even a single misplaced wall can silentl...

📖 Read original article


321. A Scalable Pipeline for LLM-Teacher Distillation Labeling: Work-Stealing Job Scheduling and Memory-Aware GPU Concurrency ​

Author: Ravi Satya Durga Prasad Yenugula
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.CL, cs.LG

arXiv:2608.15975v1 Announce Type: cross Abstract: Labeling large text corpora with LLM teachers has become a practical route to training data at scale. At millions of items, hand-labeling every batch is not feasible, and two questions dominate: what label quality a teacher buys per dollar, and how t...

📖 Read original article


322. Whose Gold? Annotator-Pool Disagreement Is Large at the Item Level, and Hidden by Small Leaderboards ​

Author: Anik Jha
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.15980v1 Announce Type: cross Abstract: Preference benchmarks are built by hiring annotators, and the identity of those annotators is treated as an implementation detail. We measure what that detail buys. On the 2,885 MultiPref items where both pools are internally unanimous, so no tie-bre...

📖 Read original article


323. Learning Varying Physical Therapist-Patient Interactions for Robot-mediated Upper Limb Task-Specific Training ​

Author: Jia Quan Loh (Human Robotics Laboratory, Department of Mechanical Engineering, The University of Melbourne), Vincent Crocher (Human Robotics Laboratory, Department of Mechanical Engineering, The University of Melbourne), Marlena Klaic (Melbourne School of Health Sciences, The University of Melbourne), Denny Oetomo (Human Robotics Laboratory, Department of Mechanical Engineering, The University of Melbourne), Ying Tan (Human Robotics Laboratory, Department of Mechanical Engineering, The University of Melbourne)
Published: 8/18/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.15995v1 Announce Type: cross Abstract: Upper extremity motor function recovery is positively linked to Task-Specific Training (TST) and sufficient therapy dosage. Rehabilitation robots can increase TST dosage via controlled, repetitive treatment and free therapists to simultaneously manag...

📖 Read original article


324. Structured Prediction for Scalable Spreadsheet Table Understanding: From Cell Types to Table Ranges (Extended Version) ​

Author: Antoine Gauquier, Ioana Manolescu, Pierre Senellart
Published: 8/18/2026, 4:00:00 AM
Categories: cs.IR, cs.DB, cs.LG

arXiv:2608.16050v1 Announce Type: cross Abstract: Spreadsheets are a primary medium for publishing tabular data, yet automatically extracting structured content from them remains difficult due to heterogeneous layouts, diverse file formats, and inconsistent organizational conventions. We address two...

📖 Read original article


325. GOD: Enhancing Generalization via Deep Grafting for Sequential Recommendation ​

Author: WooJoo Kim, JunYoung Kim, JaeHyung Lim, HwanJo Yu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.16073v1 Announce Type: cross Abstract: Sequential recommenders often struggle with sparse and noisy histories, limiting generalization to unseen interactions. Knowledge distillation mitigates this by transferring dense supervision from a teacher to a student. However, most distillation me...

📖 Read original article


326. Eigenanalysis framework for autoregressive neural emulators of multi-scale chaotic dynamics ​

Author: Conrad Ainslie, Pedram Hassanzadeh, Michael W. Mahoney, Ashesh Chattopadhyay
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, nlin.CD, physics.comp-ph

arXiv:2608.16084v1 Announce Type: cross Abstract: Neural autoregressive models have rapidly emerged as powerful emulators of high-dimensional chaotic systems, yet their long-term instability and error growth remain poorly understood, leading to ad-hoc solutions. Here, we develop an eigenanalysis fra...

📖 Read original article


327. Representation Is Not Enough: Body-Localized Thermal Evidence for Contactless Stress and Craving Sensing in Opioid Use Disorder ​

Author: Sachin Deb, Harshit Sharma, Asif Salekin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.16087v1 Announce Type: cross Abstract: Removing wearables from physiological monitoring also removes their supervision: the signal indicating where and when a stress response occurred. Contactless stress sensing therefore becomes a weakly supervised evidence-localization problem, where a ...

📖 Read original article


328. Protein Structure Prediction: From Evolutionary Constraints to Generative Modeling ​

Author: Wengan He, Yongsheng Luo, Lihong Jiang, Wenhui Xu, Yu Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.16094v1 Announce Type: cross Abstract: Accurate protein structure prediction is fundamental to structural biology because protein structure underlies molecular function and provides a basis for mechanistic interpretation. Recent advances in deep learning have transformed the field from mu...

📖 Read original article


329. EMS Coreset: An Efficient Expectation-Maximization Algorithm for Sinkhorn Coreset ​

Author: Haoyun Yin, Chuanhui Liu, Xiao Wang
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.16101v1 Announce Type: cross Abstract: Coresets distill large datasets into small, representative subsets for efficient downstream learning. Yet Optimal Transport (OT)-based selection typically requires intensive computation of transport plans, limiting scalability. We introduce a scalabl...

📖 Read original article


330. TokenSTFormer: A Tokenized Spatial-temporal Attention Model for Holistic Motion Analysis in Adolescent Idiopathic Scoliosis Screening ​

Author: Dong Chen, Kenneth M. C. Cheung
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.16122v1 Announce Type: cross Abstract: Adolescent Idiopathic Scoliosis (AIS) is a prevalent spinal deformity in adolescents that, if left untreated, can result in severe health outcomes. Traditional screening methods are limited by subjective interpretation, reliance on professional exper...

📖 Read original article


331. Coded Hankel Polynomial Chaos: Spectral Identification of Dominant Polynomial-Chaos Modes ​

Author: Zhiliang Deng, Xiaomei Yang
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.16126v1 Announce Type: cross Abstract: Identification of dominant polynomial-chaos modes is usually formulated as a sparse-regression problem on a sampled multivariate polynomial dictionary. We develop coded Hankel polynomial chaos (CH-PC), a complementary spectral formulation for dominan...

📖 Read original article


332. When Single-Dataset Conclusions Fail: A 45-Task Study of Threshold Tuning and Resampling for Imbalanced Classification ​

Author: Diyorbek Musaev
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.16147v1 Announce Type: cross Abstract: Class-imbalance handling is routinely evaluated on a single benchmark dataset, and the resulting conclusions are reported as if they were properties of the method. We show this practice is unsafe. On the public Kaggle credit-card fraud dataset, under...

📖 Read original article


333. Asymptotics-guided learning and symbolic regression for dispersive resonances ​

Author: Konstantinos Alexopoulos, Josselin Garnier
Published: 8/18/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math-ph, math.AP, math.MP, physics.optics

arXiv:2608.16152v1 Announce Type: cross Abstract: We study resonance prediction in dispersive media, formulated as nonlinear spectral problems for volume integral operators. The main idea is to use asymptotic analysis not only as a baseline approximation, but also as a guide for constructing predict...

📖 Read original article


334. Digital Twin Degradation: Detecting Cyber Physical Attacks via Temporal Inconsistencies ​

Author: Konstantinos E. Kampourakis, Vasileios Gkioulos, Sokratis Katsikas
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.16159v1 Announce Type: cross Abstract: Digital Twins (DTs) are increasingly used to monitor and analyze Cyber Physical Systems (CPS). However, in adversarial environments, the fidelity of a DT cannot be assumed. Communication delays, data manipulation, sensor degradation, or partial infor...

📖 Read original article


335. Domain-Specific Text Embedding Models for Entity Resolution ​

Author: Khajesh Sapram, Srivardhani Raju, Kishore Konda
Published: 8/18/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2608.16161v1 Announce Type: cross Abstract: General-purpose text embedding models are designed to capture semantic similarity but are not optimised for distinguishing entity records that represent the same real-world business or person. This limitation affects applications such as entity resol...

📖 Read original article


336. RadioVIL: Anomaly-Aware Diffusion Models for Radio Map Inpainting and Zero-Shot Vehicle Localization ​

Author: Ruixin Zhao, Xiucheng Wang, Qiming Zhang, Nan Cheng, Ruijin Sun, Conghao Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.16167v1 Announce Type: cross Abstract: High-precision radio map construction is essential for emerging 6G Integrated Sensing and Communication (ISAC) applications, including digital twins and intelligent transportation. However, existing deep learning methods predominantly treat this as a...

📖 Read original article


337. Decorrelation Is Not Complementarity: Skill, Not Lineage, Governs Trusted-Monitor Ensembles ​

Author: Anik Jha
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.16190v1 Announce Type: cross Abstract: Trusted monitoring has a cheap, trusted model score a stronger untrusted model's actions, and a diverse ensemble of them beats a single stronger monitor at matched cost. They are built by minimising average pairwise correlation, and that paper's twel...

📖 Read original article


338. Convolution-Free Holistic Multivariance Decomposition Layer for Efficient Hyperspectral Image Classification Tensor Networks ​

Author: S"uha Tuna, "Ulker Ba\c{s}ar
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, math.OC

arXiv:2608.16241v1 Announce Type: cross Abstract: Feature extraction for hyperspectral image classification is conventionally addressed using rigid tensor decompositions that fail to capture complex spatio-spectral interdependencies, or heavily parameterized convolutional neural networks that are co...

📖 Read original article


339. CoM$^3$eT: A foundation model for medical image analysis through federated, multidimensional context integration ​

Author: J. Raphael Sch"afer, Kai Geissler, Till Nicke, Chiara Tappermann, Karoline Heber, Eike Petersen, Habib Mergan, Lars Ole Schwen, Nick Weiss, Annika Gerken, Jan Hendrik Moltz, Tom Bisson, Isil Dogan O, Tim-Rasmus Kiehl, Norman Zerbe, Sefer Elezkurtaj, Robin S. Mayer, Nadine Flinner, Peter Wild, Isabel Dahm, Felix Peisen, Heinrich von Busch, Robert Grimm, Sebastian Arndt, Lisa Siegler, Matthias Stefan May, Antje Prasse, Natalia Artysh, Fabian Kiessling, Johannes Lotz
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.16268v1 Announce Type: cross Abstract: Medical foundation models improve generalization when training AI models with limited labeled data, but remain confined to a single specialty, such as pathology or radiology, and to either sparse or dense outputs, such as classification or segmentati...

📖 Read original article


340. Domain-Agnostic Neural Topic Modeling with Contextual Token-Level Semantic Graph Representation ​

Author: Seung-Won Seo, Won Ik Cho, Yongmin Yoo
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.16269v1 Announce Type: cross Abstract: Recent advances in neural topic models with pre-trained language models (PLMs) have achieved strong performance by leveraging general-domain pre-training, yet their topic interpretability often degrades on specialized corpora. This limitation primari...

📖 Read original article


341. Correlation Clustering with Random Partial Information ​

Author: Rajath Rao K. N., Jens Schl"oter, Sami Davies, Amira Ouchene, Yasamin Nazari
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DS, cs.LG

arXiv:2608.16315v1 Announce Type: cross Abstract: Correlation clustering is a fundamental unsupervised learning problem. On complete graphs, both the min-disagreement and min-max objectives admit constant-factor approximations, yet on general (non-complete) graphs, the best guarantees blow up to $O(...

📖 Read original article


342. Predicting, Evaluating, and Explaining Top Misinformation Spreaders via Archetypal User Behavior ​

Author: Enrico Verdolotti, Luca Luceri, Silvia Giordano
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SI, cs.CY, cs.LG

arXiv:2608.16323v1 Announce Type: cross Abstract: The spread of misinformation on social networks poses a significant challenge to online communities and society at large. Not all users contribute equally to this phenomenon: a small number of highly effective individuals can exert outsized influence...

📖 Read original article


343. LaGSplat: Inferring Physics-Governed Interactive Simulation from Monocular Video Using Latent Lagrangian Gaussian Splatting ​

Author: Louen Pottier
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.16324v1 Announce Type: cross Abstract: We present LaGSplat (Latent Lagrangian Gaussian Splatting), a framework that infers interactive, physics-governed dynamics from one or a few monocular videos. At inference it lets a user push on the filmed object, rigid or deformable, with an externa...

📖 Read original article


344. Beyond Binary Priorities: Multi-Tier SLA Scheduling for Large Language Model Serving ​

Author: Anders Vestrum, Arya Raeesi, Hanna Roed
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AR, cs.DC, cs.LG

arXiv:2608.16336v1 Announce Type: cross Abstract: Modern LLM serving deployments must simultaneously satisfy heterogeneous service-level objectives (SLOs) across a diverse population of user tiers, ranging from latency-critical API calls to background batch processing. Llumnix introduced a dynamic, ...

📖 Read original article


345. LiD-GLM: Lipschitz-constrained Deep Generalized Linear Models ​

Author: Tom Splittgerber, Niklas Koenen, Marvin N. Wright, Werner Brannath
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.16340v1 Announce Type: cross Abstract: The combination of traditional statistical models and neural network (NN) components into semi-structured hybrid models is an intriguing approach to construct models that, ideally, combine traditional interpretability with the unprecedented flexibili...

📖 Read original article


346. Architecture-Dependent Causal Transfer of Activation States Across Large Language Models ​

Author: Fernando Cardenas Piepereit
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.16347v1 Announce Type: cross Abstract: Direct communication between AI systems relies on natural language as an intermediate layer, incurring encoding/decoding overhead, token cost, and latency. We ask whether internal activation states can instead be transferred causally between differen...

📖 Read original article


347. Self-Routed Tensor Adapters for Parameter-Efficient Universal Visual Adaptation ​

Author: Suraj Yadav
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.16384v1 Announce Type: cross Abstract: Universal visual representations require adaptation mechanisms that adapt across heterogeneous domains without fragmenting knowledge into domain-specific modules. Parameter-efficient fine-tuning adapts frozen visual foundation models efficiently, but...

📖 Read original article


348. Mint-Agent: Introducing Finance-Native Agentic Foundation Models ​

Author: Agent Team, B. Zhang, Yaze Geng, Lei Tang, Yaoyang Yi, Zonghan Wu, Yifan Hu, Kun Wang, Qingsong Wen, Yilei Shao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.16386v1 Announce Type: cross Abstract: Financial agents must do more than recall domain knowledge: they must be both reliable, executing precise operations over grounded evidence, and executive, sustaining long-horizon research whose conclusions remain auditable. We present Mint-Agent, a ...

📖 Read original article


349. POI Recommendation with LLM-Augmented Multi-Graph Learning and Contrastive Alignment ​

Author: Burak Tamer, Wolfram H"opken, Zehui Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.16407v1 Announce Type: cross Abstract: Point-of-interest (POI) recommendation models based on graph neural networks achieve strong performance by propagating collaborative signals over user-item interactions, yet they struggle with the cold-start problem, where items with few or no intera...

📖 Read original article


350. Self-Supervised Noise2Noise-Enhanced Denoising for Continuous-Scan Air-Plasma THz Spectroscopy ​

Author: Adam Umra, Oways Alsoloh, Oliver Nagy, Aydin Sezgin, Clara Saraceno
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.16454v1 Announce Type: cross Abstract: Terahertz time-domain spectroscopy (THz-TDS) based on air-plasma generation and balanced air-biased coherent detection offers gap-free broadband coverage, but individual continuous-scan traces are strongly affected by pulse-to-pulse fluctuations and ...

📖 Read original article


351. Offline Reinforcement Learning for Hemodynamic Management of Sepsis in the ICU: a MIMIC-IV Study with Dual Off-Policy Evaluation ​

Author: Marc P'erez-Roig, David Fern'andez-Narro, Carlos S'aez
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.16482v1 Announce Type: cross Abstract: The dosing of intravenous fluids and vasopressors in sepsis is a sequential decision made under uncertainty and guided largely by clinical judgment, which makes it a natural target for reinforcement learning from historical care. Because a learned po...

📖 Read original article


352. Improved Regret Analysis for Parallel Gaussian Process Bandit Optimization ​

Author: Shion Takeno, Shogo Iwazaki
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.16492v1 Announce Type: cross Abstract: This paper studies the regret analysis for parallel Gaussian process (GP) bandit optimization. The known regret upper bounds for the widely used GP batched upper confidence bound and GP batched Thompson sampling (GP-BTS) suffer from a multiplicative ...

📖 Read original article


353. Density-Reweighted Entropic Optimal Transport: Decoupling Geometry from Sampling Density ​

Author: Keyi Li, Yuval Kluger, Boris Landa
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP, stat.ME

arXiv:2608.16506v1 Announce Type: cross Abstract: Dataset alignment is a central step in data analysis across science and engineering, where the goal is to match observations between datasets. Entropic Optimal Transport (EOT) offers a computationally tractable framework for this task by encoding cro...

📖 Read original article


354. LLMs for Zero-Shot Threat Detection via Structured Risk Indicators ​

Author: Abdullah Alghamdi, Siamak Layeghy, Marius Portmann
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.NI

arXiv:2608.16508v1 Announce Type: cross Abstract: We propose a two-stage large language model (LLM) framework for zero-shot detection of insider threats and advanced persistent threats (APTs) from heterogeneous security logs. The framework models user activity as chronological timelines and incorpor...

📖 Read original article


355. Data-Driven Reconstruction of Spatially Resolved Electron and Ion Energy Distributions from Macroscopic Plasma Quantities with Deep Neural Networks ​

Author: Libin Varghese, Kaushik Prajapati, Bhaskar Chaudhury
Published: 8/18/2026, 4:00:00 AM
Categories: physics.plasm-ph, cs.LG, physics.comp-ph

arXiv:2608.16519v1 Announce Type: cross Abstract: Spatially resolved EEDFs/IEDFs provide essential kinetic information about low-temperature plasmas (LTPs) and play a central role in determining transport, chemical reaction rates, and plasma surface interactions. While kinetic simulations directly r...

📖 Read original article


356. Automating Learner Assessment: Benchmarking Machine Learning and Deep Learning Models for EEG-Based Familiarity Prediction ​

Author: Isuru Nanayakkara, Thilina Halloluwa
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.HC, cs.LG

arXiv:2608.16541v1 Announce Type: cross Abstract: Objective assessment of learning remains a fundamental challenge in education. Electroencephalography (EEG) provides a direct, non-invasive window into the neural correlates of knowledge acquisition, including cognitive familiarity. This study benchm...

📖 Read original article


357. Supervising the Path to Fine Scales: GalerkinFlow for Scientific-Field and Image Super-Resolution ​

Author: Zikang Zhan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.CE, cs.LG

arXiv:2608.16546v1 Announce Type: cross Abstract: Most super-resolution models learn from paired data by supervising only the final high-resolution output. This provides little control over how the prediction should evolve between the downsampled observation and its fine target. We introduce Galerki...

📖 Read original article


358. Probabilistic Circuits as Reasoning Machines in Artificial Intelligence (Part I) ​

Author: Robert Peharz
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, math.PR

arXiv:2608.16565v1 Announce Type: cross Abstract: This cumulative habilitation thesis studies probabilistic circuits (PCs) as a powerful and tractable framework for reasoning and learning under uncertainty in artificial intelligence (AI). It first advocates for probability as a core language for AI,...

📖 Read original article


359. Degradation-Aligned Self-Supervised Learning for State of Health Estimation of Lithium-Ion Batteries under Label Sparsity ​

Author: Jiaqi Yao, Julia Kowal
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG

arXiv:2608.16612v1 Announce Type: cross Abstract: An accurate estimation of the state of health (SOH) underpins a safe and optimized use of the battery system. Although compelling, data-driven SOH estimation models typically require large amounts of high-quality labeled cycling data, while in practi...

📖 Read original article


360. The Working Set of a Coding Agent: Coherence Debt in Repository-Scale Tasks ​

Author: Bardia Mohammadi, Lars Klein, Aman Chadha, Akhil Arora, Laurent Bindschaedler
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2608.16630v1 Announce Type: cross Abstract: Repository-scale coding requires an agent to keep tests, imports, configuration, and migration rules consistent within a bounded context window. We model this as reconstructing a coupled-fact graph: at each edit, a required fact comes from recent con...

📖 Read original article


361. Toward Better Assessment of LLMs' Performance in Clinical Error Detection ​

Author: Yifan Zhang, Rahmatollah Beheshti
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.16643v1 Announce Type: cross Abstract: Automated detection of errors in clinical documentation is a promising application of large language models (LLMs), yet decisions to deploy such models rest on benchmarks that evaluate each clinical note in isolation. Error-detection benchmarks are t...

📖 Read original article


362. Turning spectra into images improves plant trait retrieval with 2D-CNNs ​

Author: Javier Lopatin, Teja Kattenborn, Eya Cherif, Sebasti'an Moreno
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.16661v1 Announce Type: cross Abstract: Hyperspectral reflectance spectroscopy enables non-destructive estimation of plant functional traits, yet current deep learning approaches process spectra as one-dimensional sequences, which limits how they capture long-range inter-band dependencies....

📖 Read original article


363. Random Quadratic Form with random forcing: Metastable synchronization by noise ​

Author: Anna Shalova
Published: 8/18/2026, 4:00:00 AM
Categories: math.PR, cs.LG, math.DS

arXiv:2608.16664v1 Announce Type: cross Abstract: We study the Random Quadratic Form (RQF) on a sphere in the presence of random Brownian forcing. We show that the forcing does not effectively change the law of the process but affects the synchronization properties of the system. While the RQF witho...

📖 Read original article


364. Hide&Seek: Learning to Explain in an End-to-End Differentiable Network ​

Author: Tal Ellinson, Hadi Mohasel Afshar, Sally Cripps
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.16689v1 Announce Type: cross Abstract: Instance-wise feature selection is a valuable tool for interpreting labeled data and the predictions of black-box models. In contrast to global feature selection techniques, instance-wise methods dynamically identify important features for each insta...

📖 Read original article


365. Learning to Price with Persuasion ​

Author: Maria-Florina Balcan, Tejas Pagare, Karan Singh
Published: 8/18/2026, 4:00:00 AM
Categories: cs.GT, cs.LG, econ.TH

arXiv:2608.16699v1 Announce Type: cross Abstract: Motivated by modern marketplaces, where the platform or the seller routinely gathers detailed user profiles, we study a novel learning theoretic model that simultaneously involves information and mechanism design. Specifically, we consider the econom...

📖 Read original article


366. MIRROR: Multimodal Intelligent Radiology Reasoning and Observation Reporter ​

Author: Vignesh Nagarajan, Sriram Venkatapathy
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.16709v1 Announce Type: cross Abstract: A radiologist reading a model's output faces two problems. The model returns a number and no reason, and any system that turns that number into readable prose can quietly add claims the model never made. MIRROR is a research prototype built to separa...

📖 Read original article


367. ClawGym II: Exploring Black-Box RL on Agent Harness ​

Author: Huatong Song, Fei Bai, Ming Yang, Renyuan Li, Jia Deng, Jujie He, Zhange Zhang, Daixuan Cheng, Yan Xing, Qi Yun, Xuxing Chen, Danyang Li, Feng Chang, Chuan Hao, Ran Tao, Jian Yang, Bryan Dai, Wayne Xin Zhao, Mingjie Tang, Ji-Rong Wen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.16798v1 Announce Type: cross Abstract: Agent harnesses have substantially improved performance on long-horizon tasks by coordinating agent interactions with the environment. However, reinforcement learning through complex harnesses remains largely unexplored, as scaling such training to l...

📖 Read original article


368. Unsupervised Learning of Cell Instances with Generative Routing Pyramids ​

Author: Ziwen Liu, Martin Weigert
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, q-bio.QM

arXiv:2608.16810v1 Announce Type: cross Abstract: Identifying and representing object instances such as cells or nuclei is a common task in microscopy image analysis. Established machine learning workflows typically use supervised detection or segmentation followed by feature extraction or classific...

📖 Read original article


369. zLend: A Dual-Scope Cash-Flow Reconstruction Framework for On-Chain Credit Underwriting ​

Author: Girish G N, Ashutosh Sahoo, Akshay SP, Gurukiran S, Dhanashekar Kandaswamy
Published: 8/18/2026, 4:00:00 AM
Categories: q-fin.RM, cs.LG

arXiv:2608.16856v1 Announce Type: cross Abstract: Decentralized lending lacks a credit bureau: a borrower's capacity to repay must be inferred entirely from public on-chain activity, without income verification or a liability record. This paper presents zLend, a deployed cash-flow underwriting frame...

📖 Read original article


370. The canonical facets of multi-separator polytopes ​

Author: Bjoern Andres, Silvia Di Gregorio, Jannik Irmai, Lucas Fabian Naumann, Shengxian Zhao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DM, cs.LG, math.CO

arXiv:2608.16861v1 Announce Type: cross Abstract: We initiate a polyhedral study of the graph multi-separator problem proposed by Irmai et al. (2024) as an alternative to the lifted multicut problem for application to the task of image segmentation. Starting with an integer linear program (ILP) form...

📖 Read original article


371. Non-Crossing Deep Quantile Regression for Distributional Survival Prediction ​

Author: Shuai Huang, Zhe Qu, Zhaowei Hua, Guohao Shen, Rui Tang, Hongtu Zhu
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP

arXiv:2608.16864v1 Announce Type: cross Abstract: In survival analysis the way covariates act on the risk of an event often differs between early and late failure times, yet hazard- and mean-based summaries collapse this variation into a single number. Quantile-based modeling instead describes the f...

📖 Read original article


372. AutoSR: Automatic Symbolic Regression by Searching Research States ​

Author: Kejia Zhang, Youran Sun, Xinyu Ren, Chugang Yi, Haizhao Yang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SC, cs.AI, cs.LG, cs.NA, math.NA

arXiv:2608.16876v1 Announce Type: cross Abstract: We introduce Automatic Symbolic Regression (AutoSR), a fully automated system that instantiates Research-Space Symbolic Regression by searching persistent scientific investigations rather than isolated equations. Finite, noisy data often yield numeri...

📖 Read original article


373. Spectral Gaps of Hit-and-Run and Coordinate Hit-and-Run ​

Author: Yunbum Kook, Santosh S. Vempala
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, math.PR, math.ST, stat.TH

arXiv:2608.16878v1 Announce Type: cross Abstract: For any convex body $\mathcal{K}\subset\mathbb{R}^{n}$ containing a unit ball, the spectral gap of Hit-and-Run is $\Omega(1/(n^2 C_{\mathsf{PI}}))$, where $C_{\mathsf{PI}}$ is the Poincar'e constant of the uniform distribution $\pi$ over $\mathcal{K...

📖 Read original article


374. Improving the matrix multiplication exponent with modern optimization and AlphaEvolve ​

Author: Emilien Dupont, Marvin Eisenberger, Borislav Kozlovskii, Abbas Mehrabian, Francisco J. R. Ruiz, Abigail See, Renfei Zhou, Josh Alman, Virginia Vassilevska Williams, Matej Balog
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DS, cs.AI, cs.CC, cs.LG

arXiv:2608.16884v1 Announce Type: cross Abstract: The current best bounds on the matrix multiplication exponent $\omega$ are obtained through a refinement of the laser method called combination loss analysis (Duan et al., 2022; Williams et al., 2024; Alman et al., 2025). In this note, we address the...

📖 Read original article


375. Sequential Batch Learning in Finite-Action Linear Contextual Bandits ​

Author: Yanjun Han, Zhengqing Zhou, Zihao Hu, Jose Blanchet, Peter W. Glynn, Yinyu Ye, Zhengyuan Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, stat.ML

arXiv:2004.06321v2 Announce Type: replace Abstract: We study the sequential batch learning problem in linear contextual bandits with finite action sets, where the decision maker is constrained to split incoming individuals into (at most) a fixed number of batches and can only observe outcomes for th...

📖 Read original article


376. Pointer Networks with Q-Learning for Combinatorial Optimization ​

Author: Alessandro Barro
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2311.02629v5 Announce Type: replace Abstract: We introduce the Pointer Q-Network (PQN), a hybrid neural architecture that integrates model-free Q-value policy approximation with Pointer Networks (Ptr-Nets) to enhance the optimality of attention-based sequence generation, focusing on long-term ...

📖 Read original article


377. DirMixE: Harnessing Test Agnostic Long-tail Recognition with Hierarchical Label Variations ​

Author: Zhiyong Yang, Qianqian Xu, Sicong Li, Zitai Wang, Xiaochun Cao, Qingming Huang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2405.07780v3 Announce Type: replace Abstract: This paper explores test-agnostic long-tail recognition, a challenging long-tail task where the test label distributions are unknown and arbitrarily imbalanced. We argue that the variation in these distributions can be broken down hierarchically in...

📖 Read original article


378. Balancing Optimality and Diversity: Human-Centered Decision Making through Generative Curation ​

Author: Michael Lingzhi Li, Shixiang Zhu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, math.OC

arXiv:2409.11535v3 Announce Type: replace Abstract: Many decision-support systems recommend actions by optimizing measurable objectives, even when a human decision-maker retains final authority and considers additional criteria that are difficult to specify in advance. We study how an algorithm shou...

📖 Read original article


379. Limits to scalable evaluation at the frontier: LLM as Judge won't beat twice the data ​

Author: Florian E. Dorner, Vivian Y. Nastl, Moritz Hardt
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2410.13341v4 Announce Type: replace Abstract: High quality annotations are increasingly a bottleneck in the explosively growing machine learning ecosystem. Scalable evaluation methods that avoid costly annotation have therefore become an important research ambition. Many hope to use strong exi...

📖 Read original article


380. MoE-Enhanced Explainable Deep Manifold Transformation for Complex Data Embedding and Visualization ​

Author: Zelin Zang, Yuhao Wang, Jinlin Wu, Hong Liu, Yue Shen, Zhen Lei, Stan Z. Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2410.19504v3 Announce Type: replace Abstract: Dimensionality reduction (DR) plays a crucial role in various fields, including data engineering and visualization, by simplifying complex datasets while retaining essential information. However, achieving both high DR accuracy and strong explainab...

📖 Read original article


381. Model Inversion Attacks: A Survey of Approaches and Countermeasures ​

Author: Zhanke Zhou, Jianing Zhu, Fengfei Yu, Xuan Li, Xiong Peng, Tongliang Liu, Bo Han
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2411.10023v3 Announce Type: replace Abstract: Deep neural networks have enabled numerous studies and applications on both Euclidean data, such as images and text, and non-Euclidean data, such as graphs. Because these networks may process private data, their deployment raises concerns about pri...

📖 Read original article


382. Rethinking Token-wise Feature Caching: Accelerating Diffusion Transformers with Dual Feature Caching ​

Author: Chang Zou, Shikang Zheng, Evelyn Zhang, Runlin Guo, Haohang Xu, Zhengyi Shi, Conghui He, Xuming Hu, Linfeng Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2412.18911v3 Announce Type: replace Abstract: Diffusion Transformers (DiT) have become the dominant methods in image and video generation yet still suffer substantial computational costs. As an effective approach for DiT acceleration, feature caching methods are designed to cache the features ...

📖 Read original article


383. ConfRetro: a 3D-aware template-free method for enhancing retrosynthesis via molecular conformer information ​

Author: Jiaxi Zhuang, Yu Zhang, Ying Qian, Aimin Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2501.12434v3 Announce Type: replace Abstract: Motivation: Retrosynthesis plays a crucial role in organic synthesis and drug discovery, focusing on identifying a set of reactants capable of synthesizing a target product molecule. Although the existing approaches have shown promising results, th...

📖 Read original article


384. Towards Unified Approaches in Self-Supervised Event Stream Modeling: Progress and Prospects ​

Author: Levente Z'olyomi, Tianze Wang, Sofiane Ennadir, Oleg Smirnov, Lele Cao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2502.04899v3 Announce Type: replace Abstract: The proliferation of digital interactions across diverse domains, such as healthcare, e-commerce, gaming, and finance, has resulted in the generation of vast volumes of event stream (ES) data. ES data comprises continuous sequences of timestamped e...

📖 Read original article


385. Leveraging Machine Unlearning for Cost-Efficient Preference Alignment ​

Author: Xiaohua Feng, Yuyuan Li, Huwei Ji, Jiaming Zhang, Li Zhang, Tianyu Du, Chaochao Chen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2504.06659v2 Announce Type: replace Abstract: Despite advances in Preference Alignment (PA) for Large Language Models (LLMs), mainstream methods like reinforcement learning with human feedback face notable challenges. These approaches require high-quality datasets of positive preference exampl...

📖 Read original article


386. A Generative Deep Learning Workflow for Inverse Molecular Design of Fuels ​

Author: Kiran K. Yalamanchi, Pinaki Pal, Balaji Mohan, Abdullah S. AlRamadan, Jihad A. Badra, Yuanjiang Pei
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph

arXiv:2504.12075v4 Announce Type: replace Abstract: In the present work, a generative deep learning framework combining a Co-optimized Variational Autoencoder (Co-VAE) with quantitative structure-property relationship (QSPR) techniques is developed to enable inverse molecular design of fuels. The Co...

📖 Read original article


387. Hybrid quantum recurrent neural network for remaining useful life prediction of turbofan engines ​

Author: Olga Tsurkan, Aleksandra Konstantinova, Arsenii Senokosov, Asel Sagingalieva, Alexey Melnikov
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, quant-ph

arXiv:2504.20823v3 Announce Type: replace Abstract: Accurate remaining useful life (RUL) estimation underpins safe operation and cost-effective maintenance of aerospace propulsion systems. We propose a Hybrid Quantum Recurrent Neural Network (HQRNN) for jet-engine RUL forecasting on the NASA C-MAPSS...

📖 Read original article


388. FDA-Opt: Federated Fine-Tuning via Dynamic Update Schedules ​

Author: Michael Theologitis, Vasilis Samoladas, Antonios Deligiannakis
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2505.04535v4 Announce Type: replace Abstract: Federated Learning (FL) enables the utilization of vast, previously inaccessible data sources. At the same time, pre-trained Language Models (LMs) have taken the world by storm and for good reason. They exhibit remarkable emergent abilities and are...

📖 Read original article


389. WATCH: Adaptive Monitoring for AI Deployments via Weighted-Conformal Martingales ​

Author: Drew Prinster, Xing Han, Anqi Liu, Suchi Saria
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2505.04608v5 Announce Type: replace Abstract: Responsibly deploying artificial intelligence (AI) / machine learning (ML) systems in high-stakes settings arguably requires not only proof of system reliability, but also continual, post-deployment monitoring to quickly detect and address any unsa...

📖 Read original article


390. One-shot Robust Federated Learning of Independent Component Analysis ​

Author: Dian Jin, Xin Bing, Yuqian Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2505.20532v2 Announce Type: replace Abstract: This paper studies robust one-shot aggregation for distributed and federated Independent Component Analysis (ICA). In this setting, each client computes a local ICA estimator, while the server aims to recover a common global mixing matrix without a...

📖 Read original article


391. PhyxMamba: Chaotic System Reconstruction from Short Context Observations with Generative State-Space Models ​

Author: Chang Liu, Bohao Zhao, Jingtao Ding, Huandong Wang, Yong Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2505.23863v3 Announce Type: replace Abstract: Understanding chaotic dynamics is a fundamental problem across scientific disciplines, including climate science, neuroscience, and fluid dynamics, yet direct experimentation and intervention in such systems are often infeasible. Chaotic system rec...

📖 Read original article


392. Feature-Aware (Hyper)graph Generation via Next-Scale Prediction ​

Author: Dorian Gailhard, Enzo Tartaglione, Lirida Naviner, Jhony H. Giraldo
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.DM

arXiv:2506.01467v4 Announce Type: replace Abstract: Graph generative models perform well on small-scale structured data but struggle to scale to large, complex structures. Hierarchical approaches improve scalability but often ignore node and edge features, which are critical in real-world applicatio...

📖 Read original article


393. VirnyFlow: Optimizing ML Pipelines for Accuracy, Fairness, and Stability at Scale ​

Author: Denys Herasymuk, Anastasiia Mozghova, Nazar Protsiv, Vladyslav Sydorak, Julia Stoyanovich
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY

arXiv:2506.01584v2 Announce Type: replace Abstract: Developing machine learning (ML) systems for real-world deployment requires navigating context-dependent trade-offs among accuracy, fairness, stability, and other objectives. Existing AutoML frameworks optimize pipelines efficiently, but they fix t...

📖 Read original article


394. Contraction-Aware Reinforcement Learning for Nonlinear Control with Statistical Robustness ​

Author: Minjae Cho, Hiroyasu Tsukamoto, Huy T. Tran
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2506.15700v2 Announce Type: replace Abstract: Control contraction metrics (CCMs)-defined by Riemannian metrics under which a closed-loop system is incrementally exponentially stable-offer a constructive framework for synthesizing contracting policies in nonlinear path-tracking problems. Howeve...

📖 Read original article


395. Leveraging Generative Artificial Intelligence for Causal Inference with Unstructured Data ​

Author: Kosuke Imai, Kentaro Nakamura
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2507.03897v4 Announce Type: replace Abstract: We introduce GenAI-Powered Inference (GPI), a statistical framework for both causal and predictive inference using unstructured data, including text and images. GPI leverages open-source Generative Artificial Intelligence (GenAI) models---such as l...

📖 Read original article


396. ROC-n-reroll: How verifier imperfection affects test-time scaling ​

Author: Florian E. Dorner, Yatong Chen, Andr'e F. Cruz, Fanny Yang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2507.12399v3 Announce Type: replace Abstract: Test-time scaling aims to improve language model performance by leveraging additional compute during inference. Many works have empirically studied techniques such as Best-of-N (BoN) and Rejection Sampling (RS) that make use of a verifier to enable...

📖 Read original article


397. Generative Model Unlearning: A Survey through Target Events, Unlearning Operators, and Evaluation Protocols ​

Author: Xiaohua Feng, Jiaming Zhang, Fengyuan Yu, Chengye Wang, Li Zhang, Kaixiang Li, Yuyuan Li, Lingjuan Lyu, Chaochao Chen, Jianwei Yin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2507.19894v2 Announce Type: replace Abstract: With the rapid advancement of generative models, privacy, copyright, safety, and reliability risks have attracted growing attention. To mitigate these risks, machine unlearning has been increasingly adapted from traditional classification models to...

📖 Read original article


398. Learning Multi-Timescale Interventions under Safety and Resource Constraints ​

Author: David Mguni, Wanrong Yang, Jing Dong, Jing Peng, Ziquan Liu, Muhammad Salman Haleem, Baoxiang Wang, Dominik Wojtczak
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2508.03875v4 Announce Type: replace Abstract: Many sequential decision problems offer qualitatively different ways of influencing the environment: some interventions act immediately, whereas others induce persistent effects that continue to shape future states long after the decision that init...

📖 Read original article


399. ProteoKnight: Convolution-based Phage Virion Protein Classification and Uncertainty Analysis ​

Author: Samiha Afaf Neha, Md. Ishrak Khan, Abir Ahammed Bhuiyan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2508.07345v2 Announce Type: replace Abstract: \textbf{Introduction:} Accurate prediction of Phage Virion Proteins (PVP) is essential for genomic studies due to their crucial role as structural elements in bacteriophages. Computational tools, particularly machine learning, have emerged for anno...

📖 Read original article


400. Enhancing Differentially Private Linear Regression via Public Second-Moment ​

Author: Zilong Cao (The School of Mathematics, Northwest University), Hai Zhang (The School of Mathematics, Northwest University)
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2508.18037v2 Announce Type: replace Abstract: Leveraging information from public data has become increasingly crucial in enhancing the utility of differentially private (DP) methods. Traditional DP approaches often require adding noise based solely on private data, which can significantly degr...

📖 Read original article


401. Near-Optimal Sample Complexity Bounds for Constrained Average-Reward MDPs ​

Author: Yukuan Wei, Xudong Li, Lin F. Yang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2509.16586v2 Announce Type: replace Abstract: Recent advances have significantly improved our understanding of the sample complexity of learning in average-reward Markov decision processes (AMDPs) under the generative model. However, much less is known about the constrained average-reward MDP ...

📖 Read original article


402. Knowing When to Defer: Selective Prediction for Responsible Knowledge Tracing ​

Author: Joshua Mitton, Prarthana Bhattacharyya, Ralph Abboud, Simon Woodhead
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2509.21514v4 Announce Type: replace Abstract: Research on Knowledge Tracing (KT) models traditionally focuses on improving predictive accuracy. However, responsible real-world deployment requires models to know when to defer uncertain predictions to a human teacher. We introduce an intrinsic s...

📖 Read original article


403. Federated Self-Supervised Modulation Classification under Non-IID and Imbalanced Data ​

Author: Usman Akram, Yiyue Chen, Haris Vikalo
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP

arXiv:2510.04927v2 Announce Type: replace Abstract: Automatic modulation classification (AMC) is a core enabler of cognitive wireless systems, providing spectrum awareness and supporting adaptive communication at the network edge. However, training AMC models on centrally aggregated data incurs high...

📖 Read original article


404. Time-Correlated Video Bridge Matching ​

Author: Viacheslav Vasilev, Arseny Ivanov, Nikita Gushchin, Maria Kovaleva, Alexander Korotin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.12453v3 Announce Type: replace Abstract: Diffusion models excel in noise-to-data generation tasks, providing a mapping from a Gaussian distribution to a more complex data distribution. However, they struggle to model translations between complex distributions, limiting their effectiveness...

📖 Read original article


405. Near-Equilibrium Propagation training in nonlinear wave systems ​

Author: Karol Sajnok, Micha{\l} Matuszewski
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.quant-gas, math-ph, math.MP, physics.optics, quant-ph

arXiv:2510.16084v3 Announce Type: replace Abstract: Backpropagation learning algorithm, the workhorse of modern artificial intelligence, is notoriously difficult to implement in physical neural networks. Equilibrium Propagation (EP) is an alternative with comparable efficiency and strong potential f...

📖 Read original article


406. Explainable Heterogeneous Anomaly Detection in Financial Networks via Adaptive Expert Routing ​

Author: Zan Li, Rui Fan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE

arXiv:2510.17088v3 Announce Type: replace Abstract: Financial anomalies arise from heterogeneous mechanisms - price shocks, liquidity freezes, contagion cascades, and momentum reversals - yet existing detectors produce uniform anomaly scores without revealing which mechanism is failing or where risk...

📖 Read original article


407. Last-Iterate Analyses of FTRL with the 1/2-Tsallis Entropy in Stochastic Bandits ​

Author: Jingxin Zhan, Yuze Han, Zhihua Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.22819v3 Announce Type: replace Abstract: The convergence analysis of online learning algorithms is central to machine learning theory, where the last-iterate convergence is particularly important, as it captures the learner's actual decisions and describes the evolution of the learning pr...

📖 Read original article


408. Retrofit: Continual Learning with Controlled Forgetting for Binary Security Detection and Analysis ​

Author: Yiling He, Junchi Lei, Hongyu She, Shuo Shao, Xinran Zheng, Yiping Liu, Zhan Qin, Lorenzo Cavallaro
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2511.11439v3 Announce Type: replace Abstract: Binary security has increasingly relied on deep learning to reason about malware behavior and program semantics. However, the performance often degrades as threat landscapes evolve and code representations shift. While continual learning (CL) offer...

📖 Read original article


409. Q-Regularized Generative Auto-Bidding: From Suboptimal Trajectories to Optimal Policies ​

Author: Mingming Zhang, Na Li, Zhuang Feiqing, Hongyang Zheng, Jiangbing Zhou, Wang Wuyin, Sheng-jie Sun, XiaoWei Chen, Junxiong Zhu, Lixin Zou, Chenliang Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR

arXiv:2601.02754v3 Announce Type: replace Abstract: With the rapid development of e-commerce, auto-bidding has become a key asset in optimizing advertising performance under diverse advertiser environments. The current approaches focus on reinforcement learning (RL) and generative models. These effo...

📖 Read original article


410. Replacing Tunable Parameters in Weather and Climate Models with State-Dependent Functions using Reinforcement Learning ​

Author: Pritthijit Nath, Sebastian Schemm, Henry Moss, Peter Haynes, Emily Shuckburgh, Mark J. Webb
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2601.04268v4 Announce Type: replace Abstract: Weather and climate models rely on parametrisations to represent unresolved sub-grid processes. Traditional schemes rely on fixed coefficients that are weakly constrained and tuned offline, contributing to persistent biases that limit their ability...

📖 Read original article


411. Stitch the Fragments: One-Shot Hierarchical Federated Clustering ​

Author: Shenghong Cai, Zihua Yang, Yang Lu, Mengke Li, Yuzhu Ji, Yiqun Zhang, Yiu-Ming Cheung
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.06404v2 Announce Type: replace Abstract: Federated Clustering (FC) faces a critical bottleneck in real-world scenarios, i.e., global clusters are rarely intact, often fragmenting into incomplete, multi-granular unlabeled ``clusterlets'' distributed across Non-IID clients. Although hierarc...

📖 Read original article


412. In-Context Source and Channel Coding ​

Author: Ziqiong Wang, Tianqi Ren, Rongpeng Li, Zhifeng Zhao, Honggang Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.10267v2 Announce Type: replace Abstract: Separate Source-Channel Coding (SSCC) remains attractive for text transmission due to its modularity and compatibility with mature entropy coders and powerful channel codes. However, SSCC often suffers from a pronounced cliff effect in low Signal-t...

📖 Read original article


413. Robust Privacy: Inference-Stage Privacy through Certified Robustness ​

Author: Jiankai Jin, Xiangzheng Zhang, Zhao Liu, Wenzhuo Xu, Dongdong Yang, Deyue Zhang, Quanchen Zou
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2601.17360v3 Announce Type: replace Abstract: An adversary observing a model's released prediction can infer sensitive attributes of the queried input, or even reconstruct representatives of the model's training data. The inference interface thus acts as a side channel for privacy leakage. We ...

📖 Read original article


414. Analytical Provisioning for Attention-FFN Disaggregated LLM Serving under Stochastic Workloads ​

Author: Chendong Song, Meixuan Wang, Hang Zhou, Hong Liang, Yuan Lyu, Zixi Chen, Yuwei Fan, Zijie Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.21351v4 Announce Type: replace Abstract: Attentio-FFN disaggregation (AFD) is an emerging architecture for LLM decoding that separates state-heavy, KV-cache-dominated Attention computation from stateless, compute-intensive FFN computation, connected by per-step communication. While AFD en...

📖 Read original article


415. Evaluating the trustworthiness of the Fr\'echet Inception Distance with stochastic embedding representations ​

Author: Ciaran Bench, Vivek Desai, Carlijn Roozemond, Ruben van Engen, Spencer A. Thomas
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.21979v2 Announce Type: replace Abstract: Feature embeddings acquired from pretrained models are widely used in medical applications of deep learning to assess the characteristics of datasets; e.g. to determine the quality of synthetic, generated medical images. The Fr'{e}chet Inception D...

📖 Read original article


416. PRO-Bid: Pareto-Prioritized Regret Optimization for Constraint-Aware Generative Auto-Bidding ​

Author: Binglin Wu, Yingyi Zhang, Xianneng Li, Ruyue Deng, Chuan Yue, Weiru Zhang, Xiaoyi Zeng
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.GT

arXiv:2602.08261v2 Announce Type: replace Abstract: Auto-bidding systems strive to maximize marketing value while maintaining high compliance with efficiency constraints, such as Target Cost-Per-Action (CPA). While Decision Transformers offer powerful sequence modeling capabilities, their applicatio...

📖 Read original article


417. Grounding LTL Tasks in Sub-Symbolic RL Environments for Zero-Shot Generalization ​

Author: Matteo Pannacci, Andrea Fanti, Elena Umili, Roberto Capobianco
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.09761v2 Announce Type: replace Abstract: In this work we address the problem of training a Reinforcement Learning agent to follow multiple temporally-extended instructions expressed in Linear Temporal Logic in sub-symbolic environments. Previous multi-task work has mostly relied on knowle...

📖 Read original article


418. Adaptive Optimization via Momentum on Variance-Normalized Gradients ​

Author: Francisco Patitucci, Aryan Mokhtari
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2602.10204v2 Announce Type: replace Abstract: We introduce MVN-Grad (Momentum on Variance-Normalized Gradients), an Adam-style optimizer that improves stability and performance by combining two complementary ideas: variance-based normalization and momentum applied after normalization. MVN-Grad...

📖 Read original article


419. Closing the Loop: A Control-Theoretic Framework for Provably Stable Time Series Forecasting with LLMs ​

Author: Xingyu Zhang, Jingyao Wang, Zeen Song, Changwen Zheng, Wenwen Qiang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.12756v2 Announce Type: replace Abstract: Large Language Models (LLMs) have recently shown exceptional potential in time series forecasting (TSF), leveraging their inherent sequential reasoning capabilities to model complex temporal dynamics. Existing approaches typically employ an autoreg...

📖 Read original article


420. Zero-Shot Instruction Following in RL via Structured LTL Representations ​

Author: Mathias Jackermeier, Mattia Giuri, Jacques Cloete, Alessandro Abate
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.14344v2 Announce Type: replace Abstract: We study instruction following in multi-task reinforcement learning, where an agent must zero-shot execute novel tasks not seen during training. In this setting, linear temporal logic (LTL) has recently been adopted as a powerful framework for spec...

📖 Read original article


421. MDP Planning as Policy Inference ​

Author: David Tolpin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.17375v3 Announce Type: replace Abstract: We formulate episodic Markov decision process (MDP) planning as Bayesian inference over policies. The primary contribution is conceptual: the policy itself is treated as the latent variable, and expected return defines an unnormalized posterior den...

📖 Read original article


422. LORA-CRAFT: Cross-layer Rank Adaptation via Frozen Tucker Decomposition of Pre-trained Attention Weights ​

Author: Kasun Dewage, Marianna Pensky, Suranadi De Silva, Shankadeep Mondal
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17510v2 Announce Type: replace Abstract: We introduce LoRA-CRAFT (\textbf{C}ross-layer \textbf{R}ank \textbf{A}daptation via \textbf{F}rozen \textbf{T}ucker), abbreviated CRAFT throughout, an extremely parameter-efficient fine-tuning (PEFT) method that applies Tucker tensor decomposition ...

📖 Read original article


423. Measuring the Prevalence of Policy Violating Content with ML Assisted Sampling and LLM Labeling ​

Author: Attila Dobi, Aravindh Manickavasagam, Benjamin Thompson, Xiaohan Yang, Faisal Farooq
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2602.18518v2 Announce Type: replace Abstract: Content safety teams need metrics that reflect what users actually experience, not only what is reported. We study prevalence: the fraction of user views (impressions) that went to content violating a given policy on a given day. Accurate prevalenc...

📖 Read original article


424. Exact Attention Sensitivity and the Geometry of Transformer Stability ​

Author: Seyed Morteza Emadi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.18849v2 Announce Type: replace Abstract: We develop a sensitivity analysis for transformer attention in a geometry aligned with tokenwise computation. Our main result is the exact identity $|J_\tau(u)|{\infty\to1}=\theta(p)/\tau$ for the Jacobian $J\tau(u)$ of the tempered softmax $u...

📖 Read original article


425. Decentralized Federated Learning by Partial Message Exchange ​

Author: Shan Sha, Shenglong Zhou, Xin Wang, Lingchen Kong, Geoffrey Ye Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.01730v4 Announce Type: replace Abstract: Decentralized federated learning (DFL) has emerged as a transformative server-free paradigm that enables collaborative learning over large-scale heterogeneous networks. However, it continues to face fundamental challenges, including data heterogene...

📖 Read original article


426. Joint MDPs and Reinforcement Learning in Coupled-Dynamics Environments ​

Author: Ege C. Kaya, Mahsa Ghasemi, Abolfazl Hashemi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2603.06946v2 Announce Type: replace Abstract: Many distributional quantities in reinforcement learning are intrinsically joint across actions, including distributions of gaps and probabilities of superiority. However, the classical Markov decision process (MDP) formalism specifies only margina...

📖 Read original article


427. Not All Neighbors Matter: Understanding the Impact of Graph Sparsification on GNN Pipelines ​

Author: Yuhang Song, Naima Abrar Shami, Romaric Duvignau, Vasiliki Kalavri
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.DB

arXiv:2603.06952v2 Announce Type: replace Abstract: As graphs scale to billions of nodes and edges, graph Machine Learning workloads are constrained by the cost of multi-hop traversals over exponentially growing neighborhoods. While various system-level and algorithmic optimizations have been propos...

📖 Read original article


428. Differentiable Thermodynamic Phase-Equilibria for Machine Learning ​

Author: Karim K. Ben Hicham, Moreno Ascani, Jan G. Rittig, Alexander Mitsos
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.11249v4 Announce Type: replace Abstract: Accurate prediction of phase equilibria remains a central challenge in chemical engineering. Physics-consistent machine learning methods that incorporate thermodynamic structure into neural networks have recently shown strong performance for activi...

📖 Read original article


429. Informative Perturbation Selection for Uncertainty-Aware Post-hoc Explanations ​

Author: Sumedha Chugh, Ranjitha Prasad, Nazreen Shah
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2603.14894v3 Announce Type: replace Abstract: Trust and ethical concerns due to the widespread deployment of opaque machine learning (ML) models motivating the need for reliable model explanations. Post-hoc model-agnostic explanation methods addresses this challenge by learning a surrogate mod...

📖 Read original article


430. Free-Flow Class-Incremental Learning: Towards Robust CIL under Variable Class Arrivals ​

Author: Zhiming Xu, Baile Xu, Jian Zhao, Furao Shen, Suorong Yang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.02765v2 Announce Type: replace Abstract: Class-incremental learning (CIL) is commonly evaluated under predefined schedules with fixed or nearly equal class increments, leaving irregular class-arrival scenarios underexplored. However, practical CIL systems may need to update whenever new c...

📖 Read original article


431. A Theoretical Framework for Statistical Evaluability of Generative Models ​

Author: Shashaank Aiyer, Yishay Mansour, Shay Moran, Han Shao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT

arXiv:2604.05324v3 Announce Type: replace Abstract: Statistical evaluation aims to estimate the generalization performance of a model using held-out i.i.d. test data sampled from the ground-truth distribution. In supervised learning settings such as classification, performance metrics such as error ...

📖 Read original article


432. From Uniform to Learned Knots: A Study of Spline-Based Numerical Encodings for Tabular Deep Learning ​

Author: Manish Kumar, Anton Frederik Thielmann, Christoph Weisser, Benjamin S"afken
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.05635v2 Announce Type: replace Abstract: Numerical preprocessing remains a critical component of tabular deep learning, as the representation of continuous features can strongly affect downstream performance. We systematically study spline-based numerical encodings, including B-splines, M...

📖 Read original article


433. CodeQuant: Unified Clustering and Quantization for Enhanced Outlier Smoothing in Low-Precision Mixture-of-Experts ​

Author: Xiangyang Yin, Xingyu Liu, Tianhua Xia, Bo Bao, Vithursan Thangarasa, Valavan Manohararajah, Eric Sather, Sai Qian Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.10496v2 Announce Type: replace Abstract: Outliers have emerged as a fundamental bottleneck in preserving accuracy for low-precision large models, particularly within Mixture-of-Experts (MoE) architectures that are increasingly central to large-scale language modeling. Under post-training ...

📖 Read original article


434. The Optimal Sample Complexity of Multiclass and List Learning ​

Author: Chirag Pabbaraju
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2604.24749v3 Announce Type: replace Abstract: While the optimal sample complexity of binary classification in terms of the VC dimension is well-established, determining the optimal sample complexity of multiclass classification has remained open. The appropriate complexity parameter for multic...

📖 Read original article


435. Learning Koopman operators for coupled systems via information on governing equations of subsystems ​

Author: Tatsuya Naoi, Jun Ohkubo
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.01835v2 Announce Type: replace Abstract: Nonlinear coupled systems are ubiquitous in science and engineering. The analysis and modeling of such systems are challenging due to their high dimensionality and complex interactions among subsystems. In recent years, operator-theoretic methods b...

📖 Read original article


436. Delay, Plateau, or Collapse: Evaluating the Impact of Systematic Verification Error on RLVR ​

Author: Kazuki Egashira, Mark Vero, Jasper Dekoninck, Florian E. Dorner, Robin Staab, Martin Vechev
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.02909v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a powerful approach for improving the reasoning capabilities of large language models (LLMs). While RLVR is designed for tasks with verifiable ground-truth answers, real-world verifie...

📖 Read original article


437. Can Attribution Predict Risk? From Multi-View Attribution to Planning Risk Signals in End-to-End Autonomous Driving ​

Author: Le Yang, Haijun Liu, Jiawei Liang, ShangQuan Sun, Xiaochun Cao
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.06264v2 Announce Type: replace Abstract: End-to-end autonomous driving models generate future trajectories from multi-view inputs, improving system integration but introducing opaque decisions and hard-to-localize risks. Existing methods either rely on auxiliary monitoring models or gener...

📖 Read original article


438. Convex Optimization with Nested Evolving Feasible Sets ​

Author: Karthick Krishna M., Haricharan Balasundaram, Rahul Vaze
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, math.OC

arXiv:2605.07386v2 Announce Type: replace Abstract: \emph{Convex Optimization with Nested Evolving Feasible Sets (CONES)} is considered where the objective function (f) remains fixed but the feasible region evolves over time as a nested sequence (S_1 \supseteq S_2 \supseteq \cdots \supseteq S_T)...

📖 Read original article


439. Learning to Bid with Unknown Private Values in Budget-Constrained First-Price Auctions ​

Author: Zihao Hu, Yuxiao Wen, Yuan Yao, Jiheng Zhang, Zhengyuan Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.09448v2 Announce Type: replace Abstract: We study the operational problem of automated bidding in repeated first-price auctions under budget and return-on-spend (RoS) constraints. In this setting, an auto-bidder must translate advertiser goals and constraints into real-time bids while lea...

📖 Read original article


440. Scaling Laws for Dynamic Mini-Batch SGD in Sketched Linear Regression ​

Author: Ziyan Chen, Zhongzhu Zhou, Ding-Xuan Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.24316v3 Announce Type: replace Abstract: Mini-batching is central to large-scale optimization, yet its role in statistical scaling laws remains limited. We study one-pass and multi-pass batch SGD for sketched linear regression under power-law spectral and source conditions. Our analysis r...

📖 Read original article


441. RT-Lynx: Putting GEMM Sparsity in the Right Place for Diffusion Models ​

Author: Xing Cong, Hanlin Tang, Kan Liu, Tao Lan, Lin Qu, Chenhao Xie
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.26632v3 Announce Type: replace Abstract: Diffusion Transformers (DiT) achieve strong performance in image generation but incur substantial inference costs. While prior work has reduced this cost via quantization and distillation, semi-structured sparsity, which can nearly halve FLOPs, rem...

📖 Read original article


442. Periodic Topological Deep Learning for Polymer Design and Discovery ​

Author: Yasharth Yadav, Tze Kwang Gerald Er, Atsushi Goto, Kelin Xia
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.26833v2 Announce Type: replace Abstract: Polymers underpin applications across energy, healthcare, and materials science, yet their vast chemical space makes systematic discovery challenging. Most machine learning approaches represent polymers as molecular graphs of a single repeating uni...

📖 Read original article


443. The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level Energy Attribution ​

Author: Deepak Panigrahy, Aakash Tyagi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR, cs.DC, cs.PF

arXiv:2605.27599v3 Announce Type: replace Abstract: Agentic AI workloads - where a single user goal triggers multi-step orchestration, tool calls, retries, and failure recovery - are being targeted for edge deployment, with NVIDIA, Dell, HP, ASUS, MSI, Acer, and Gigabyte all shipping GB10-based desk...

📖 Read original article


444. Representation Signatures and Risk-Feedback Alignment in LLM Trading Agents ​

Author: Weicheng Xue
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP

arXiv:2605.28850v3 Announce Type: replace Abstract: We study behavioral alignment and representation dynamics of large language model (LLM) agents in financial decision environments. TradeArena, an auditable trading-agent testbed with risk reports, execution simulation, memory, and replayable trajec...

📖 Read original article


445. Annealed Softmax Greedy in Many-Armed Bayesian Bandits ​

Author: William Overman, Mohsen Bayati
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.31034v3 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards and group-based policy optimization methods update a stochastic policy by sampling multiple completions per prompt and increasing the policy's probability on those with higher reward. These updates, un...

📖 Read original article


446. Learning symplectic model reduction based on an approximation theorem of symplectic embeddings ​

Author: Liyi Feng, Yifa Tang, Yulin Xie, Ruili Zhang, Aiqing Zhu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.04623v2 Announce Type: replace Abstract: High-dimensional Hamiltonian systems play a central role in many scientific and engineering disciplines, with dynamics that evolve on symplectic manifolds. Although deep learning provides powerful tools for constructing its low-dimensional surrogat...

📖 Read original article


447. Provably Efficient Personalized Multi-Objective Bandits with Proactive Conversational Queries ​

Author: Linfeng Cao, Ming Shi, Ness B. Shroff
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.08410v2 Announce Type: replace Abstract: Personalized decision-making in multi-objective bandits requires learning user-specific trade-offs among competing objectives. Since arm utility depends on both unknown rewards and unknown preferences, existing methods infer preferences only from u...

📖 Read original article


448. Speculative Rollback Correction for Quality-Diverse Web Agent Imitation ​

Author: Longkun Hao, Hongyu Lin, Hao Li, Zhuowen Liu, Zhichao Yang, Haojie Hao, Dongshuo Huang, Haitao Yang, Hongyu Ge, Ming jie Xie, Yanjun Wu, Zi Hao Yin, Yan Bai, Yihang Lou
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.12485v2 Announce Type: replace Abstract: Training interactive web agents through imitation learning from expert trajectories has emerged as a highly effective approach. However, determining the optimal timing for expert intervention presents a critical challenge in this context. Delayed i...

📖 Read original article


449. Sum-of-Squares Degree Barriers for the Reweighted-Hinge Method in Robust Halfspace Learning: A Christoffel-Function Characterization ​

Author: Xiaoyu Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, stat.ML

arXiv:2606.17215v2 Announce Type: replace Abstract: A certificate that removes outliers sees the data only through its low-degree moments, and an adversary exploits exactly this, hiding corruption where the clean data already looks typical, in the blind spot no bounded-degree test resolves. That bli...

📖 Read original article


450. Spectral Certificates and Projection-DPP Rounding for Determinantal MAP Selection ​

Author: Richard Yi Da Xu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.19411v3 Announce Type: replace Abstract: Selecting a fixed-size subset that maximizes the determinant of a positive semidefinite kernel is the MAP problem for a size-constrained determinantal point process and the classical maximum-entropy sampling problem. Although this discrete problem ...

📖 Read original article


451. SL-S4Wave: Self-Supervised Learning of Physiological Waveforms with Structured State Space Models ​

Author: Feng Wu, Harsh Deep, Eric Lehman, Sanyam Kapoor, Guoshuai Zhao, Rahul G. Krishnan, Gari Clifford, Li-wei H Lehman
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.19888v2 Announce Type: replace Abstract: Modeling long-sequence medical time series data, such as electrocardiograms (ECG), poses significant challenges due to high sampling rates, multichannel signal complexity, inherent noise, and limited labeled data. While recent self-supervised learn...

📖 Read original article


452. LACE-SVD: Loss-Aware SVD with Cumulative Error Correction for LLM Compression ​

Author: Zhuowen Liu, Longkun Hao, Shiyu Feng, Xiaowen Chang, Ruiqun Li, Changqun Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.03057v2 Announce Type: replace Abstract: The rapid growth in the parameter scale of large language models (LLMs) has created a strong demand for efficient compression techniques. As a hardware-agnostic and highly compatible approach, low-rank compression has been widely adopted to reduce ...

📖 Read original article


453. Directional Curvature from Armijo Backtracking: A Low-Cost Sharpness Probe and a Calibration-Free Learning-Rate Safeguard for Adam ​

Author: Ashmitha R, J"org Frochte
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.03998v4 Announce Type: replace Abstract: The local sharpness of the loss, the top Hessian eigenvalue $\lambda_1$, determines the largest stable gradient step, but measuring it normally requires Lanczos or Hessian-vector products. A single Armijo backtracking line search already carries th...

📖 Read original article


454. Efficient Safety Alignment of Language Models via Latent Personality Traits ​

Author: Mohamed Amine Merzouk, Nolan Smyth, Damiano Fornasiere, Linh Le, David Williams-King, Adam Oberman
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CR

arXiv:2607.07918v2 Announce Type: replace Abstract: Current safety methods for large language models are known to be vulnerable to adversarial attacks, motivating research into robust alternatives. Latent Adversarial Training (LAT) is among the most effective defenses, but can degrade utility and re...

📖 Read original article


455. LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning ​

Author: Ning Liu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10139v3 Announce Type: replace Abstract: Selecting the correct answer from a pool of candidate reasoning chains is the engine of test-time scaling, yet the standard selectors each carry a cost: self-consistency inherits the errors of the single model it resamples, and trained reward model...

📖 Read original article


456. Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values ​

Author: Jan Betley, Johannes Treutlein, Jan Dubi'nski, Harry Mayne, Karol Ga{\l}\k{a}zka, Niels Warncke, Anna Sztyber-Betley, Owain Evans
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2607.14345v4 Announce Type: replace Abstract: People use language models for practical questions whose answers are difficult to verify. We show that models exhibit covert value leakage: the information they provide is influenced by their own values, without this influence being disclosed to th...

📖 Read original article


457. A Blueprint for Equilibrium-Based Differentiable Continuous-Variable Thermodynamic Computing ​

Author: Owen Lockwood, J'er'emy B'ejanin, Joost Bus, Christopher Chamberland, Patrick Huembeli, Frank Sch"afer, Guillaume Verdon
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.ET, physics.app-ph

arXiv:2607.16183v2 Announce Type: replace Abstract: To help address the escalating energy and latency demands of machine-learning workloads, we introduce a blueprint for an energy-efficient and fast thermodynamic computing stack that leverages stochastic analog processes in physical hardware. In thi...

📖 Read original article


458. SCPP: A Unified Python Library for Soft Clustering ​

Author: Kiyan Rezaee, Morteza Ziabakhsh, Artin Bahrampour, Seyed Mohammad Ghoreishi, Asal Khaje, Ali Sajedifar, Manny Chalak, Ava Zerafatangiz, Sadegh Eskandari
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19620v2 Announce Type: replace Abstract: In this paper, we present SCPP (Soft Clustering Python Package), an open-source Python framework for soft clustering. SCPP establishes a canonical, scikit-learn-compatible estimator interface that standardizes model training, prediction, membership...

📖 Read original article


459. Global Convergence of DGM and PINN Algorithms for Solving Nonlinear PDEs ​

Author: Justin Sirignano, Konstantinos Spiliopoulos, Samuel Cohen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2607.24726v2 Announce Type: replace Abstract: The Deep Galerkin Method (DGM) and Physics Informed Neural Networks (PINNs) have become widely-used methods for solving partial differential equations (PDEs) in the rapidly growing field of scientific machine learning. In these methods, a neural ne...

📖 Read original article


460. Neurai-VN Benchmark: Standardized Machine Learning Models for Multimodal Digital Phenotyping in Mental Health Classification ​

Author: Quoc-Cuong Pham, Hoang-Thuy-Duong Vu, Thi-Thanh-Huong Ha, Huy-Hieu Pham
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25232v3 Announce Type: replace Abstract: Digital phenotyping (DP) using smartphones and wearable devices has emerged as a promising approach for assessing mental health, particularly depression and anxiety. However, progress remains difficult to evaluate because of heterogeneity across da...

📖 Read original article


461. ODEWorld: A Continuous Predictive Architecture via Physical-Time Flow ​

Author: Dongxiu Liu, Haoyi Niu, Peng Cheng, Yuan Gao, Xirui Kang, Sangli Teng, Koushil Sreenath, Xianyuan Zhan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.RO

arXiv:2607.27924v3 Announce Type: replace Abstract: In the physical world we inhabit, space and time are fundamentally continuous. However, existing machine learning paradigms for world modeling are largely confined to discrete-time prediction, thereby exhibiting significant inefficiency in capturin...

📖 Read original article


462. Hierarchical Copula-Gumbel-Top-K Routing: Two-Sided Dependence Control for Frozen Mixture-of-Experts at Fixed Per-Token Routing Laws ​

Author: Richard Yi Da Xu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.28670v3 Announce Type: replace Abstract: A stochastic Gumbel-Top-K router defines, for every token of a mixture-of-experts (MoE) model, a routing law: a distribution over ordered expert lists and mixture weights. We ask which joint distributions over the routing choices of different token...

📖 Read original article


Author: Mohammad Asif, Azizuddin Khan, Mohd Azam, Anurag Rajkumar Bombarde
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET

arXiv:2607.28687v2 Announce Type: replace Abstract: As populations age, cognitive decline from mild cognitive impairment (MCI) to dementia is a defining health challenge of the coming decades, yet routine assessment often misses its earliest signs. This article critically synthesizes recent technolo...

📖 Read original article


464. Learning Optimal Dynamic Matching via Graph Neural Networks ​

Author: Genta Okada, Shunya Noda, Junpei Komiyama, Akira Matsushita
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.GT, econ.TH

arXiv:2607.28925v2 Announce Type: replace Abstract: Dynamic matching markets require decisions about whom to match and when: matching now yields value but removes participants who may create better future opportunities. We develop a value-based reinforcement-learning framework for this problem on fi...

📖 Read original article


465. A 2-Block Architecture for Real-Time EEG Gait Decoding: A Pilot Study ​

Author: Shantanu Sarkar, Saurabh Prasad, Jose L. Contreras-Vidal
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.HC, eess.SP

arXiv:2608.02083v2 Announce Type: replace Abstract: Closed-loop lower-limb exoskeleton control via Electroencephalography (EEG) remains limited by motion artifacts, low signal-to-noise ratio, and binary gait formulations that fail to capture full cortical gait complexity. We propose a 2-block Brain-...

📖 Read original article


466. Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't ​

Author: Ravi Satya Durga Prasad Yenugula
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.02829v3 Announce Type: replace Abstract: Model families are typically trained size by size, each from scratch. Can a pretrained large model instead be converted into a smaller sibling? We characterize the 1.4B->410M conversion in Pythia end to end. Representations align strongly across si...

📖 Read original article


467. A Physics-Informed Hybrid Neural Operator for Transient Magnetization Prediction in Power Magnetics ​

Author: Yachao Zhu, Qiujie Huang, Sinan Li, Yang Li, Gang Lei, Jianguo Zhu
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci, cs.SY, eess.SY

arXiv:2608.02965v2 Announce Type: replace Abstract: Magnetic components in high-frequency, high-power-density converters are increasingly driven by non-sinusoidal flux-density waveforms with fast transitions, minor-loop operation, dc bias, and temperature variation. Under these conditions, steady-st...

📖 Read original article


468. SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation ​

Author: Zikun Qu, Min Zhang, Mingze Kong, Zhiwei Shang, Zhengyu Chen, Yikun Ban, Shuang Qiu, Zhongxiang Dai
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.04419v2 Announce Type: replace Abstract: On-policy distillation (OPD) provides dense teacher supervision on student-generated trajectories, but standard reverse-KL training can assign insufficient probability to other plausible continuations. Teacher entropy alone does not reveal whether ...

📖 Read original article


469. A Probabilistic Circuit-Induced Pseudo-Metric for Out-of-Distribution Detection ​

Author: Bhumika K, Vidhya S, Narayanan C Krishnan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09117v2 Announce Type: replace Abstract: Probabilistic Circuits (PCs) are tractable generative models whose internal nodes encode a hierarchy of probabilistic summaries over different variable scopes. Existing PC-based out-of-distribution (OOD) detection methods ignore this hierarchy, red...

📖 Read original article


470. From Recoverability to Functional Use: Auditing Temporal Reports in Time-Series Forecasting ​

Author: Qipeng Qian, Yuntao Qian
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.10433v4 Announce Type: replace Abstract: Time-series forecasters increasingly accompany numerical predictions with explicit temporal reports, such as delays or selected history, but a correct report need not describe the information actually used by the forecast. We separate this problem ...

📖 Read original article


471. JAPE: Joint Anomaly Prediction and Intrinsic Explanation in Multivariate Time Series ​

Author: Yian Wei, Yuanyuan Yao, Lu Chen, Xiangmin Zhou, Tianyi Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.11801v2 Announce Type: replace Abstract: Multivariate time-series anomaly prediction aims to identify whether and when anomalies will occur over a future horizon from historical observations. Existing methods primarily characterize anomalies as deviations in future numerical values, which...

📖 Read original article


472. Represent, Then Generate: Multimodal-Conditioned Time-Series Generation under Irregular Missingness ​

Author: Haochen Zhang, Jiaheng Guo, Yu-Chao Huang, Nicholas Konz, Tianlong Chen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12592v2 Announce Type: replace Abstract: Continuous physiological time series underpin modern clinical monitoring, yet many of the most informative signals are invasive, expensive, or simply unavailable for a given patient. Conditional generation offers a remedy: an absent signal can be s...

📖 Read original article


473. From Fixed Grids to Moving Particles:A Transferable Latent Operator for Fluid Dynamics ​

Author: Meng Li, Chuqi Chen, Zhengqing Gao, Xi Zhou, Xiao Sun, Yang Xiang, Huaxi Huang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GR

arXiv:2608.14120v2 Announce Type: replace Abstract: Lagrangian modeling is vital to fluid dynamics, as it characterizes particle transport and complements the Eulerian representation. However, Lagrangian trajectories are less commonly available than Eulerian fields, while most neural operators are t...

📖 Read original article


474. From Monte Carlo to neural networks approximations of boundary value problems ​

Author: Lucian Beznea, Iulian Cimpean, Oana Lupascu-Stamate, Ionel Popescu, Arghir Zarnescu
Published: 8/18/2026, 4:00:00 AM
Categories: math.PR, cs.AI, cs.LG, cs.NA, math.AP, math.NA

arXiv:2209.01432v4 Announce Type: replace-cross Abstract: In this paper we study probabilistic and neural network approximations for solutions to Poisson equation subject to Holder data in general bounded domains of $\mathbb{R}^d$. We aim at two fundamental goals. The first, and the most important, ...

📖 Read original article


475. Doubly robust nearest neighbors in factor models ​

Author: Raaz Dwivedi, Sabina Tomkins, Predrag Klasnja, Susan Murphy, Devavrat Shah
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2211.14297v4 Announce Type: replace-cross Abstract: We introduce and analyze an improved variant of nearest neighbors (NN) for estimation with missing data in latent factor models. We consider a matrix completion problem with missing data, where the $(i, t)$-th entry, when observed, is given b...

📖 Read original article


476. EquiPocket: an E(3)-Equivariant Geometric Graph Neural Network for Ligand Binding Site Prediction ​

Author: Yang Zhang, Zhewei Wei, Ye Yuan, Chongxuan Li, Wenbing Huang
Published: 8/18/2026, 4:00:00 AM
Categories: q-bio.BM, cs.LG

arXiv:2302.12177v4 Announce Type: replace-cross Abstract: Predicting the binding sites of target proteins plays a fundamental role in drug discovery. Most existing deep-learning methods consider a protein as a 3D image by spatially clustering its atoms into voxels and then feed the voxelized protein...

📖 Read original article


477. Score Attack: A Lower Bound Technique for Optimal Differentially Private Learning ​

Author: T. Tony Cai, Yichen Wang, Linjun Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: math.ST, cs.CR, cs.LG, stat.ME, stat.ML, stat.TH

arXiv:2303.07152v3 Announce Type: replace-cross Abstract: Achieving optimal statistical performance while ensuring the privacy of personal data is a challenging yet crucial objective in modern data analysis. However, characterizing the optimality, particularly the minimax lower bound, under privacy ...

📖 Read original article


478. Dynamic Pricing and Advertising with Demand Learning ​

Author: Shipra Agrawal, Yiding Feng, Wei Tang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.GT, cs.LG

arXiv:2304.14385v4 Announce Type: replace-cross Abstract: We consider a novel pricing and advertising framework in which a seller not only sets the product price but also designs flexible advertising schemes to influence customers' valuations of the product. We impose no structural restriction on th...

📖 Read original article


479. ODTlearn: A Package for Learning Optimal Decision Trees for Prediction and Prescription ​

Author: Patrick Vossler, Nathan Justin, Sina Aghaei, Nathanael Jo, Andr'es G'omez, Phebe Vayanos
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC

arXiv:2307.15691v4 Announce Type: replace-cross Abstract: ODTlearn is an open source Python package that provides methods for learning optimal decision trees for high-stakes predictive and prescriptive tasks based on the state-of-the-art mixed-integer optimization (MIO) framework proposed in Aghaei ...

📖 Read original article


480. ShadowNet for Data-Centric Quantum System Learning ​

Author: Yuxuan Du, Yibo Yang, Tongliang Liu, Zhouchen Lin, Bernard Ghanem, Dacheng Tao
Published: 8/18/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG

arXiv:2308.11290v2 Announce Type: replace-cross Abstract: Understanding the dynamics of large quantum systems is hindered by the curse of dimensionality. Statistical learning offers new possibilities in this regime through neural network protocols and classical shadows, while both methods have limit...

📖 Read original article


481. Macroeconomic Forecasting with Large Language Models ​

Author: Andrea Carriero, Davide Pettenuzzo, Shubhranshu Shekhar
Published: 8/18/2026, 4:00:00 AM
Categories: econ.EM, cs.CL, cs.LG

arXiv:2407.00890v5 Announce Type: replace-cross Abstract: This paper presents a comparative analysis evaluating the accuracy of Large Language Models (LLMs) against traditional macro time series forecasting approaches. In recent times, LLMs have surged in popularity for forecasting due to their abil...

📖 Read original article


482. Quantum Large Language Models via Tensor Network Disentanglers ​

Author: Borja Aizpurua, Fernando Loren, Saeed S. Jahromi, Sukhbinder Singh, Roman Orus
Published: 8/18/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG

arXiv:2410.17397v2 Announce Type: replace-cross Abstract: We introduce a framework for seamlessly integrating quantum computing into pretrained large language models (LLMs). The key idea is to construct a hybrid quantum-classical representation that exactly reproduces the original model, providing a...

📖 Read original article


483. Multi-Bin Batching for Increasing LLM Inference Throughput ​

Author: Ozgur Guldogan, Jackson Kunde, Kangwook Lee, Ramtin Pedarsani
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.DC, cs.LG, cs.SY, eess.SY

arXiv:2412.04504v2 Announce Type: replace-cross Abstract: As large language models (LLMs) grow in popularity for their diverse capabilities, improving the efficiency of their inference systems has become increasingly critical. Batching LLM requests is a critical step in scheduling the inference jobs...

📖 Read original article


484. Experimentally Extending Quantum Kernel Learning to Quantum Data by NMR ​

Author: Vivek Sabarad, Vishal Varma, T. S. Mahesh
Published: 8/18/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, physics.app-ph

arXiv:2412.09557v3 Announce Type: replace-cross Abstract: Quantum kernel learning (QKL) promises efficient machine learning by encoding feature maps onto exponentially large Hilbert spaces inherent in quantum systems. Using the liquid-state nuclear magnetic resonance (NMR) platform, we implement and...

📖 Read original article


485. Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities ​

Author: Qirun Dai, Dylan Zhang, Jiaqi W. Ma, Hao Peng
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2501.12147v2 Announce Type: replace-cross Abstract: Selecting appropriate training data is crucial for instruction fine-tuning of large language models (LLMs), which aims to (1) elicit strong capabilities, and (2) achieve balanced performance across different tasks. Influence-based methods sho...

📖 Read original article


486. Thinking Outside the (Gray) Box: A Context-Based Score for Assessing Value and Originality in Neural Text Generation ​

Author: Giorgio Franceschelli, Mirco Musolesi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CY, cs.LG

arXiv:2502.13207v4 Announce Type: replace-cross Abstract: Despite the increasing use of large language models for creative tasks, their outputs often lack diversity. Common solutions, such as sampling at higher temperatures, can compromise the quality of the results. Dealing with this trade-off is s...

📖 Read original article


487. Bringing Generative Learning to Representation Learning: Self-Supervised Transfer Learning as Distribution Matching ​

Author: Yuling Jiao, Wensen Ma, Defeng Sun, Hansheng Wang, Yang Wang
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, stat.ME

arXiv:2502.14424v3 Announce Type: replace-cross Abstract: Most self-supervised learning objectives defend against collapse but leave the target representation law unspecified. We formulate representation learning as Distribution Matching (DM), learning an augmentation-invariant encoder whose induced...

📖 Read original article


488. Helios 2.0: A Robust, Ultra-Low Power Gesture Recognition System Optimised for Event-Sensor based Wearables ​

Author: Prarthana Bhattacharyya, Joshua Mitton, Ryan Page, Owen Morgan, Oliver Powell, Benjamin Menzies, Gabriel Homewood, Kemi Jacobs, Paolo Baesso, Taru Muhonen, Richard Vigars, Louis Berridge
Published: 8/18/2026, 4:00:00 AM
Categories: cs.HC, cs.CV, cs.LG

arXiv:2503.07825v3 Announce Type: replace-cross Abstract: We present an advance in wearable technology: a mobile-optimized, real-time, ultra-low-power event camera system that enables natural hand gesture control for smart glasses, dramatically improving user experience. While hand gesture recogniti...

📖 Read original article


489. Diffusion-model approach to flavor models: A case study for $S_4^\prime$ modular flavor model ​

Author: Satsuki Nishimura, Hajime Otsuka, Haruki Uchiyama
Published: 8/18/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-th

arXiv:2504.00944v2 Announce Type: replace-cross Abstract: We propose a numerical method of searching for parameters with experimental constraints in generic flavor models by utilizing diffusion models, which are classified as a type of generative artificial intelligence (generative AI). As a specifi...

📖 Read original article


490. On Stopping Times of Power-one Sequential Tests: Tight Lower and Upper Bounds ​

Author: Shubhada Agrawal, Ashwin Ram, Aaditya Ramdas
Published: 8/18/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.ML, stat.TH

arXiv:2504.19952v2 Announce Type: replace-cross Abstract: We present two general lower bounds for stopping times of sequential tests between arbitrary composite nulls $\mathcal P$ and alternatives $\mathcal Q$. The first lower bound is for the ``Wald setting'' where the type-1 error level $\alpha$ a...

📖 Read original article


491. Multilook Coherent Imaging: Theoretical Guarantees and Algorithms ​

Author: Xi Chen, Soham Jana, Christopher A. Metzler, Arian Maleki, Shirin Jalali
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, eess.IV

arXiv:2505.23594v2 Announce Type: replace-cross Abstract: Multilook coherent imaging is a widely used technique in applications such as digital holography, ultrasound imaging, and synthetic aperture radar. A central challenge in these systems is the presence of multiplicative noise, commonly known a...

📖 Read original article


492. FuseLIP: Multimodal Embeddings via Early Fusion of Discrete Tokens ​

Author: Christian Schlarmann, Francesco Croce, Nicolas Flammarion, Matthias Hein
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2506.03096v2 Announce Type: replace-cross Abstract: Contrastive language-image pre-training aligns features of text-image pairs in a common latent space via distinct encoders for each modality. While this approach achieves impressive performance in several zero-shot tasks, it cannot natively h...

📖 Read original article


493. Lightweight Deep Learning-Based Channel Estimation for RIS-Aided Extremely Large-Scale MIMO Systems on Resource-Limited Edge Devices ​

Author: Muhammad Kamran Saeed, Ashfaq Khokhar, Shakil Ahmed
Published: 8/18/2026, 4:00:00 AM
Categories: cs.IT, cs.CV, cs.LG, cs.NI, math.IT

arXiv:2507.09627v3 Announce Type: replace-cross Abstract: Next-generation wireless technologies such as 6G aim to meet demanding requirements such as ultra-high data rates, low latency, and enhanced connectivity. Extremely Large-Scale MIMO (XL-MIMO) and Reconfigurable Intelligent Surface (RIS) are k...

📖 Read original article


494. Projection-based multifidelity linear regression for data-scarce applications ​

Author: Vignesh Sella, Julie Pham, Karen Willcox, Anirban Chaudhuri
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.CE, cs.LG

arXiv:2508.08517v2 Announce Type: replace-cross Abstract: Surrogate modeling for systems with high-dimensional quantities of interest remains challenging, particularly when training data are costly to acquire. This work develops multifidelity methods for multiple-input multiple-output linear regress...

📖 Read original article


495. Encoding of musical structures in hidden units of restricted Boltzmann machines ​

Author: Mutsumi Kobayashi, Hiroshi Watanabe
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS

arXiv:2509.04899v4 Announce Type: replace-cross Abstract: Restricted Boltzmann machines (RBMs) are energy-based models originating from statistical physics, in which hidden units mediate the probability distribution of high-dimensional visible configurations. In this study, we use symbolic music as ...

📖 Read original article


496. SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA ​

Author: Haozhou Xu, Dongxia Wu, Matteo Chinazzi, Ruijia Niu, Rose Yu, Yi-An Ma
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2509.25459v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) show promise in generating long-form scientific explanations that synthesize evidence and connect multiple factors. However, in long-form scientific question answering, LLMs often hallucinate, producing unsupporte...

📖 Read original article


497. Comprehensive language-image pre-training for 3D medical image understanding ​

Author: Tassilo Wald, Ibrahim Ethem Hamamci, Yuan Gao, Sam Bond-Taylor, Harshita Sharma, Maximilian Ilse, Cynthia Lo, Olesya Melnichenko, Anton Schwaighofer, Noel C. F. Codella, Maria Teodora Wetscherek, Klaus H. Maier-Hein, Panagiotis Korfiatis, Valentina Salvatelli, Javier Alvarez-Valle, Fernando P'erez-Garc'ia
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2510.15042v3 Announce Type: replace-cross Abstract: In the 3D medical image domain, vision-language pre-training is used to create vision-language encoders (VLEs) that can support radiologists by retrieving patients with similar abnormalities, predicting likelihoods of abnormality, or, with do...

📖 Read original article


498. Iso-Riemannian Optimization on Learned Data Manifolds ​

Author: Willem Diepeveen, Melanie Weber
Published: 8/18/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DG

arXiv:2510.21033v3 Announce Type: replace-cross Abstract: We develop a theory of iso-Riemannian optimization for problems constrained to learned data manifolds, a setting in which classical Riemannian optimization - and Riemannian gradient descent in particular - can be poorly suited. That is, favor...

📖 Read original article


499. From Lab to Reality: A Practical Evaluation of Deep Learning Models and LLMs for Vulnerability Detection ​

Author: Chaomeng Lu, Bert Lagaisse
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.SE

arXiv:2512.10485v3 Announce Type: replace-cross Abstract: Vulnerability detection methods based on deep learning (DL) have shown strong performance on benchmark datasets, yet their real-world effectiveness remains underexplored. Recent work suggests that graph neural network-based and transformer-ba...

📖 Read original article


500. On Fibonacci Ensembles: An Alternative Approach to Ensemble Learning Inspired by the Timeless Architecture of the Golden Ratio ​

Author: Ernest Fokou'e
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2512.22284v2 Announce Type: replace-cross Abstract: Nature rarely reveals her secrets bluntly, yet in the Fibonacci sequence she grants us a glimpse of her quiet architecture of growth, harmony, and recursive stability \citep{Koshy2001Fibonacci, Livio2002GoldenRatio}. From spiral galaxies to t...

📖 Read original article


501. Gradient descent reliably finds depth- and gate-optimal circuits for generic unitaries ​

Author: Janani Gomathi, Alex Meiburg
Published: 8/18/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2601.03123v2 Announce Type: replace-cross Abstract: When the gate set has continuous parameters, synthesizing a unitary operator as a quantum circuit is, in principle, always possible using exact methods. However, efficiently finding depth- and gate-minimal circuits remains a major challenge. ...

📖 Read original article


502. Misconception Diagnosis From Student-Tutor Dialogue: Generate, Retrieve, Rerank ​

Author: Joshua Mitton, Prarthana Bhattacharyya, Digory Smith, Thomas Christie, Ralph Abboud, Simon Woodhead
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2602.02414v2 Announce Type: replace-cross Abstract: Timely and accurate identification of student misconceptions is key to improving learning outcomes and pre-empting the compounding of student errors. However, this task is highly dependent on the effort and intuition of the teacher. In this w...

📖 Read original article


503. MiDAS: A Multimodal Data Acquisition System and Dataset for Robot-Assisted Minimally Invasive Surgery ​

Author: Keshara Weerasinghe (MD), Seyed Hamid Reza Roodabeh (MD), Andrew Hawkins (MD), Zhaomeng Zhang, Zachary Schrader, Homa Alemzadeh
Published: 8/18/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2602.12407v3 Announce Type: replace-cross Abstract: Background: Robot-assisted minimally invasive surgery (RMIS) research increasingly relies on multimodal data, yet access to proprietary robot telemetry remains a major barrier. We introduce MiDAS, an open-source, platform-agnostic system enab...

📖 Read original article


504. ReLoop: Structured Modeling and Behavioral Verification for Reliable LLM-Based Optimization ​

Author: Junbo Jacob Lian, Yujun Sun, Huiling Chen, Chaoyu Zhang, Hanzhang Qin, Chung-Piaw Teo
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG, math.OC

arXiv:2602.15983v3 Announce Type: replace-cross Abstract: Large language models (LLMs) can translate natural language into optimization code, but silent failures pose a critical risk: code that executes and returns solver-feasible solutions may encode semantically incorrect formulations---a feasibil...

📖 Read original article


505. mmWave Radar Aware Dual-Conditioned GAN for Speech Reconstruction of Signals With Low SNR ​

Author: Jash Karani, Adithya Chittem, Deepan Roy, Sandeep Joshi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SD, cs.LG

arXiv:2602.22431v2 Announce Type: replace-cross Abstract: Millimeter-wave (mmWave) radar captures are band-limited and noisy, making for difficult reconstruction of intelligible full-bandwidth speech. In this work, we propose a two-stage speech reconstruction pipeline for mmWave using a Radar-Aware ...

📖 Read original article


506. Random Quadratic Form on a Sphere: Synchronization by Common Noise ​

Author: Maximilian Engel, Anna Shalova
Published: 8/18/2026, 4:00:00 AM
Categories: math.PR, cs.LG, math.DS

arXiv:2603.06187v2 Announce Type: replace-cross Abstract: We introduce the Random Quadratic Form (RQF): a stochastic differential equation which formally corresponds to the gradient flow of a random quadratic functional on a sphere. While the one-point dynamics of the system is a Brownian motion and...

📖 Read original article


507. Mixture of experts architectures for machine learning interatomic potentials ​

Author: Yuzhi Liu, Duo Zhang, Anyang Peng, Weinan E, Linfeng Zhang, Han Wang
Published: 8/18/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, physics.comp-ph

arXiv:2603.07977v3 Announce Type: replace-cross Abstract: Machine Learning Interatomic Potentials (MLIPs) enable accurate large-scale atomistic simulations, yet improving their expressive capacity efficiently remains challenging. Here we systematically investigate Mixture-of-Experts (MoE) and Mixtur...

📖 Read original article


508. HiAP: A Multi-Granular Stochastic Auto-Pruning Framework for Vision Transformers ​

Author: Andy Li, Aiden Durrant, Milan Markovic, Georgios Leontidis
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2603.12222v3 Announce Type: replace-cross Abstract: Vision Transformers require significant computational resources and memory bandwidth, severely limiting their deployment on resource-constraint hardware. Most structured pruning methods reduce theoretical cost effectively, yet they typically ...

📖 Read original article


509. Less Data, Faster Convergence: Goal-Driven Data Optimization for Multimodal Instruction Tuning ​

Author: Rujie Wu, Haozhe Zhao, Hai Ci, Yizhou Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2603.12478v3 Announce Type: replace-cross Abstract: Multimodal instruction tuning is often compute-inefficient because training budgets are spread across large mixed image-video pools whose utility is highly uneven. We present Goal-Driven Data Optimization (GDO), a framework that computes six ...

📖 Read original article


510. Deep Learning for BioImaging: What Are We Really Learning? ​

Author: Ivan Svatko, Maxime Sanchez, Ihab Bendidi, Gilles Cottrell, Auguste Genovesio
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2603.13377v2 Announce Type: replace-cross Abstract: Representation learning has driven major advances in natural image analysis by enabling models to acquire high-level semantic features. In microscopy imaging, however, it remains unclear what current representation learning methods really lea...

📖 Read original article


511. Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought ​

Author: Zichen Xie, Wenxi Wang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2603.18334v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) increasingly assist secure software development, their ability to meet the rigorous demands of Rust program verification remains unclear. Existing evaluations treat Rust verification as a black box, assessing m...

📖 Read original article


512. SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs ​

Author: Yadi Cao, Sicheng Lai, Jiahe Huang, Yang Zhang, Zach Lawrence, Rohan Bhakta, Izzy F. Thomas, Mingyun Cao, Chung-Hao Tsai, Zihao Zhou, Yidong Zhao, Hao Liu, Alessandro Marinoni, Alexey Arefiev, Rose Yu
Published: 8/18/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.AI, cs.DC, cs.LG

arXiv:2603.20253v4 Announce Type: replace-cross Abstract: Evaluating LLM agents for scientific tasks has focused on token costs while ignoring tool-use costs like simulation time and experimental resources. As a result, metrics like pass@k become impractical under realistic budget constraints. To ad...

📖 Read original article


513. Camera-Agnostic Pruning of 3D Gaussian Splats via Descriptor-Based Beta Evidence ​

Author: Peter Fasogbon, Ugurcan Budak, Patrice Rondao Alface, Hamed Rezazadegan Tavakoli
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2603.21933v2 Announce Type: replace-cross Abstract: The pruning of 3D Gaussian splats is essential for reducing their complexity to enable efficient storage, transmission, and downstream processing. However, most of the existing pruning strategies depend on camera parameters, rendered images, ...

📖 Read original article


514. FactReview: Evidence-Grounded Peer Review with Execution-Based Claim Verification ​

Author: Ling Yue, Chaoqian Ouyang, Hang Xu, Ruijun Huang, Yuchen Liu, Libin Zheng, Wei Liu, Shaowu Pan, Shimin Di, Min-Ling Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2604.04074v4 Announce Type: replace-cross Abstract: Large language model (LLM)-based reviewing systems typically assess manuscripts in isolation, leaving literature- and code-dependent claims difficult to verify. We present FactReview, an audit pipeline that extracts review-relevant claims, gr...

📖 Read original article


515. Deterministic Adam-Inspired Methods with Accelerated Convergence Rate ​

Author: Yaxin Yu, Long Chen, Zeyi Xu
Published: 8/18/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2604.08742v2 Announce Type: replace-cross Abstract: Adam is widely used, but its convergence theory remains incomplete even in the deterministic full-batch setting because momentum and adaptive preconditioning are tightly coupled. For smooth convex objectives, we split the momentum variable th...

📖 Read original article


516. Morphology-Conditioned World Model for Cross-Embodiment Quadrupedal Locomotion ​

Author: Mohamad H. Danesh, Chenhao Li, Amin Abyaneh, Anas Houssaini, Kirsty Ellis, Glen Berseth, Marco Hutter, Hsiu-Chin Lin
Published: 8/18/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2604.08780v2 Announce Type: replace-cross Abstract: World models promise a paradigm shift in robotics, where an agent learns the physics of its environment once and then acquires behaviors efficiently. Yet the learned dynamics models at their core are typically morphology locked. In legged loc...

📖 Read original article


517. Transferable FB-GNN-MBE Framework for Potential Energy Surfaces: Data-Adaptive Transfer Learning in Deep Learned Many-Body Expansion Theory ​

Author: Siqi Chen, Zhiqiang Wang, Yili Shen, Xianqi Deng, Xi Cheng, Cheng-Wei Ju, Jun Yi, Guo Ling, Dieaa Alhmoud, Hui Guan, Zhou Lin
Published: 8/18/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG

arXiv:2604.09320v5 Announce Type: replace-cross Abstract: Mechanistic understanding and rational design of complex chemical systems depend on fast and accurate predictions of electronic structures beyond individual building blocks. However, if the system exceeds hundreds of atoms, first-principles q...

📖 Read original article


518. Differentially Private Verification of Distribution Properties ​

Author: Elbert Du, Cynthia Dwork, Pranay Tankala, Linjun Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.DS, cs.CC, cs.LG

arXiv:2604.10819v2 Announce Type: replace-cross Abstract: A recent line of work initiated by Chiesa and Gur and further developed by Herman and Rothblum investigates the sample and communication complexity of verifying properties of distributions with the assistance of a powerful, knowledgeable, but...

📖 Read original article


519. MoRFI: Monotonic Sparse Autoencoder Feature Identification ​

Author: Dimitris Dimakopoulos, Shay B. Cohen, Ioannis Konstas
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2604.26866v2 Announce Type: replace-cross Abstract: Large language models (LLMs) acquire most of their factual knowledge during the pre-training stage, through next token prediction. Subsequent stages of post-training often introduce new facts outwith the parametric knowledge, giving rise to h...

📖 Read original article


520. RETO: A Rotary-Enhanced Transformer Operator for High-Fidelity Prediction of Automotive Aerodynamics ​

Author: Bojun Zhang, Huiyu Yang, Yunpeng Wang, Yuntian Chen, Yuanwei Bin, Rikui Zhang, Jianchun Wang
Published: 8/18/2026, 4:00:00 AM
Categories: eess.IV, cs.LG

arXiv:2605.00062v3 Announce Type: replace-cross Abstract: Rapid aerodynamic evaluation is crucial for modern vehicle design, yet existing neural operators struggle to capture intricate spatial correlations. We propose the rotary-enhanced transformer operator (RETO), a novel neural solver featuring a...

📖 Read original article


521. Leakage-Audited Benchmarking Reveals Limited Evidence for Cross-Subject Auditory-Evoked EEG Vowel Perception Decoding ​

Author: Xiaoyang Li, Zeyan Tao
Published: 8/18/2026, 4:00:00 AM
Categories: eess.SP, cs.CL, cs.CV, cs.LG, cs.SD, q-bio.NC

arXiv:2605.00865v3 Announce Type: replace-cross Abstract: We tested whether auditory-evoked EEG supports subject-independent five-vowel perception decoding when trial identity, model identity, prediction provenance, and participant-level inference are controlled within a single benchmark. We reconst...

📖 Read original article


522. Evolving Ensemble of Agents ​

Author: Zongmin Yu, Liu Yang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG

arXiv:2605.09018v4 Announce Type: replace-cross Abstract: We introduce the Evolving Ensemble of Agents (EvE), a decentralized framework that organizes existing, highly capable coding agents into a live, co-evolving system for algorithmic discovery. Rather than reinventing the wheel within the ``LLMs...

📖 Read original article


523. The Observable Wasserstein Distance ​

Author: Edivaldo Lopes dos Santos, Leandro Vicente Mauri, Washington Mio, Tom Needham
Published: 8/18/2026, 4:00:00 AM
Categories: math.MG, cs.LG

arXiv:2605.09916v2 Announce Type: replace-cross Abstract: We introduce the observable Wasserstein distance, a framework for deriving lower bounds on the Wasserstein distance between probability measures on Polish metric spaces, designed to bypass the computational intractability of exact optimal tra...

📖 Read original article


524. AgentMV: A State-Guided Multi-Agent Framework for Budget-Aware Music Video Generation ​

Author: Huimin Wang, Chang Xia, Leilei Ouyang, Yongqi Kang, Yu Fu, Yuqi Ouyang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MA

arXiv:2605.10723v2 Announce Type: replace-cross Abstract: Generating a complete music video from a song requires more than synthesizing visually plausible clips for individual lyric prompts. A practical system must maintain long-range visual consistency, coordinate recurring motifs, synchronize edit...

📖 Read original article


525. An Agentic AI Framework with Large Language Models and Chain-of-Thought for UAV-Assisted Logistics Scheduling with Mobile Edge Computing ​

Author: Hanwen Zhang, Dusit Niyato, Wei Zhang, Xin Lou, Malcolm Yoke Hean Low
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2605.13221v2 Announce Type: replace-cross Abstract: In cloud manufacturing, unmanned aerial vehicles (UAVs) can support both product collection and mobile edge computing (MEC). This joint operation forms a hybrid scheduling problem, where physical logistics decisions are coupled with computati...

📖 Read original article


526. Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting ​

Author: Amritansh Maurya, Navjot Singh, Mohammed Javed, Omar Moured
Published: 8/18/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.CV, cs.LG

arXiv:2605.20254v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown promising results on NLP tasks, however, their performance on tabular data still needs research attention, because Table Question-Answering (TQA) requires precise cell retrieval and multi-step structure...

📖 Read original article


527. The Economics of Model Collapse: Equilibrium, Welfare, and Optimal Provenance Subsidies in Synthetic Data Markets ​

Author: Gustav Olaf Yunus Laitinen-Fredriksson Lundstr"om-Imanov
Published: 8/18/2026, 4:00:00 AM
Categories: econ.GN, cs.CY, cs.LG, q-fin.EC

arXiv:2605.20279v2 Announce Type: replace-cross Abstract: Generative artificial intelligence is rapidly transforming the supply side of training data: an increasing share of new tokens, images, and structured records is produced by previous-generation models rather than by human originators. Recursi...

📖 Read original article


528. The Economics of AI Inference: Inflation Dynamics, Welfare Costs, and Optimal Monetary Policy under the Inference-Cost Phillips Curve ​

Author: Gustav Olaf Yunus Laitinen-Fredriksson Lundstr"om-Imanov
Published: 8/18/2026, 4:00:00 AM
Categories: econ.GN, cs.LG, q-fin.EC

arXiv:2605.20281v2 Announce Type: replace-cross Abstract: We develop a unified microeconomic and monetary theory of artificial intelligence inference costs and their pass-through to inflation, welfare, and optimal monetary policy. We introduce the Inference-Cost Phillips Curve (ICPC), an augmented N...

📖 Read original article


529. Practical and Optimal Algorithm for Linear Contextual Bandits with Rare Parameter Updates ​

Author: Sanghoon Yu, Min-hwan Oh
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2606.00984v2 Announce Type: replace-cross Abstract: We study linear contextual bandits under rare parameter updates: the learner may incorporate reward feedback into its parameter estimate only at a small number of update times, while still observing contexts online and selecting actions seque...

📖 Read original article


530. A Quantitative Approximation Framework for Flow Distillation in Diffusion Models ​

Author: Weiguo Gao, Ming Li, Lei Shi, Hanfei Zhou
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2606.03820v2 Announce Type: replace-cross Abstract: We develop a quantitative framework for diffusion distillation by viewing few step sampling as approximation through compositions of learned flow maps. For trajectory distillation of the probability flow ODE, we show that low noise multimodal...

📖 Read original article


531. Reward hacking in physical reinforcement learning revealed by turbulent drag reduction ​

Author: Giorgio Maria Cavallazzi, Miguel P'erez-Cuadrado, Alfredo Pinelli
Published: 8/18/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG

arXiv:2606.06227v3 Announce Type: replace-cross Abstract: Reinforcement-learning controllers optimise specified rewards, but in physical systems those rewards often capture only part of the true control objective. Three mechanisms through which this mismatch can produce apparent success without phys...

📖 Read original article


532. Computationally tractable robust differentially private mean estimation ​

Author: Kelly Ramsay
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML

arXiv:2606.12654v2 Announce Type: replace-cross Abstract: We develop a new, differentially private mean estimator called the balloon mean. The main features of the balloon mean are that it is computationally tractable and enjoys robustness to outlying observations. It is based on an iterative clippi...

📖 Read original article


533. Interpreting "Interpretability" and Explaining "Explainability" in Machine Learning in Physics ​

Author: Rikab Gambhir, Luisa Lucie-Smith, Jesse Thaler
Published: 8/18/2026, 4:00:00 AM
Categories: physics.data-an, astro-ph.GA, cs.LG, hep-ph

arXiv:2606.26228v2 Announce Type: replace-cross Abstract: We review the concepts of interpretability and explainability as they apply to machine learning in physics. We define interpretability as concerning the structural transparency of a model (the ability to understand or approximate its inner wo...

📖 Read original article


534. DanceOPD: On-Policy Generative Field Distillation ​

Author: Wei Zhou, Xiongwei Zhu, Zelin Xu, Bo Dong, Lixue Gong, Yongyuan Liang, Meng Chu, Leigang Qu, Lingdong Kong, Wei Liu, Tat-Seng Chua
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2606.27377v3 Announce Type: replace-cross Abstract: Modern image generation demands a single model that unifies diverse capabilities, including text-to-image (T2I), local editing, and global editing. However, these capabilities are rarely naturally aligned and often conflict. For instance, edi...

📖 Read original article


535. Sampling the Schwinger Model with Gauge-Equivariant Diffusion ​

Author: Octavio Vega, Aida X. El-Khadra
Published: 8/18/2026, 4:00:00 AM
Categories: hep-lat, cond-mat.str-el, cs.LG

arXiv:2606.27481v2 Announce Type: replace-cross Abstract: We present a first study of a diffusion-based approach to accelerated sampling of the $N_f = 2$ lattice Schwinger model. Our work is inspired by recent and growing successes in developing such generative models for ensemble generation in LFT ...

📖 Read original article


536. On the Role of Directionality in Structural Generalization ​

Author: Zichao Wei
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.02307v2 Announce Type: replace-cross Abstract: Several SLOG test categories explicitly involve directional distinctions (modifier position shifts, argument extraction positions), yet AM-Parser, the previous SOTA, uses an AM algebra whose operations do not encode direction. We redesign the...

📖 Read original article


537. Statistical Adversaries: Natural Backdoor-like Adversarial Features in Clean Vision Datasets ​

Author: Paul K. Mandal, Pavan Reddy, Tristan Malatynski
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CR, cs.LG

arXiv:2607.05516v2 Announce Type: replace-cross Abstract: Model-specific adversarial attacks have been extensively studied. We study a different failure mode: naturally occurring statistical signals in vision data that can behave as backdoor-like triggers without being maliciously inserted. We call ...

📖 Read original article


538. Lesioned Multimodal Language Models Reproduce Aphasic Picture-Naming Patterns ​

Author: Yong Yang, Xiang Guan, Sophie Arheix-Parras, Saeed Ahmadi, Roger Newman-Norlund, Leonardo Bonilha, Christopher Rorden, Julius Fridriksson, Rutvik H. Desai, Srihari Nelakuditi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.11621v2 Announce Type: replace-cross Abstract: Aphasia following stroke commonly produces systematic naming errors with characteristic profiles, but whether general-purpose language models not designed for clinical simulation can reproduce these patterns remains untested. We investigated ...

📖 Read original article


539. The Many Senses of Visual Similarity: A Text-Prompted Image Perceptual Metric ​

Author: Sheng-Yu Wang, Yotam Nitzan, Aaron Hertzmann, Jun-Yan Zhu, Eli Shechtman, Alexei A. Efros, Richard Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.18237v2 Announce Type: replace-cross Abstract: Human visual similarity judgments are context-dependent. For example, two images may be similar in shape but distinct in color. Existing perceptual similarity metrics, however, collapse these nuances into a single scalar value, offering no me...

📖 Read original article


540. Priors learned from legacy reconstructions inherit undetectable overconfidence ​

Author: Ali Siahkoohi, Sina Alemohammad
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.21721v3 Announce Type: replace-cross Abstract: Where truths are scarce (e.g., seismic and medical imaging), a prior for an ill-posed inverse problem is trained on an archive of legacy reconstructions---an older method's outputs---and its uncertainty is treated as data-driven. In the popul...

📖 Read original article


541. Fast Trainable Multilinear Bases for Image Compression ​

Author: Shiwen An, Zhongyi Ni, Huanhai Zhou, Jin-Guo Liu
Published: 8/18/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, math.OC, quant-ph

arXiv:2608.00053v2 Announce Type: replace-cross Abstract: The Discrete Fourier Transform, the Discrete Cosine Transform, and their block-wise variants underpin most deployed image and video codecs. Their effectiveness rests on three properties: they run in near-linear time (linear up to a polylogari...

📖 Read original article


542. UOT-IR: Structured Routing of High-Polyphony Symbolic Music into Fixed-Budget Representations ​

Author: Ziyue Kang, Nan Nan, Chenhao Lin, Xiaohong Guan
Published: 8/18/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG

arXiv:2608.00576v2 Announce Type: replace-cross Abstract: High-polyphony symbolic music is increasingly used in generation, analysis, and arrangement, yet many downstream tasks require bounded representations with fixed tracks or slots. Converting richly orchestrated scores into compact forms is the...

📖 Read original article


543. How Benchmarks and Evaluation Protocols Shape Conclusions in Provenance-Based Intrusion Detection ​

Author: Lorenzo Guerra, Thomas Chapuis, Guillaume Duc, Pavlo Mozharovskyi, Van-Tam Nguyen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.01454v2 Announce Type: replace-cross Abstract: Provenance-based intrusion detection systems (PIDS) frequently report strong performance, but the conclusions drawn from these results can be highly sensitive to benchmarking choices and evaluation protocols. We investigate this dependency by...

📖 Read original article


544. Non-KKT Accumulation in Entropic Mirror Descent ​

Author: Kuangyu Ding, Kim-Chuan Toh
Published: 8/18/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DS

arXiv:2608.01658v2 Announce Type: replace-cross Abstract: For mirror descent generated by a Legendre kernel, perhaps one of the most basic question in optimization is this: must every accumulation point of a bounded mirror descent sequence be Karush--Kuhn--Tucker (KKT) stationary under proper stepsi...

📖 Read original article


545. Integrating spectral and morphological plant features with decision-tree models for early-season cotton biomass and nitrogen status estimation from multi-year UAV data ​

Author: Vaishali Swaminathan, Nithya Rajan, J Alex Thomasson, Amrit Shrestha, Karem Meza Capcha, Robert Hardin, Pramod Pokhrel
Published: 8/18/2026, 4:00:00 AM
Categories: eess.IV, cs.LG

arXiv:2608.07801v2 Announce Type: replace-cross Abstract: Precision nitrogen (N) management (PNM) for cotton requires in-season monitoring of crop growth parameters and N status indicators to decide fertilizer timing, placement, and application rates for optimal canopy development and yield. This st...

📖 Read original article


546. The Spectral Neuron ​

Author: Alex Shtoff
Published: 8/18/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.08003v2 Announce Type: replace-cross Abstract: As machine learned models increase in complexity and expressive power, features of simpler models, such as intrinsic coefficient transparency and control over the shape of the modeled function are lost. On the one edge of the spectrum we have...

📖 Read original article


547. Population-Scalable Multi-Agent World Modeling ​

Author: Renjie Zhao, Yuxiang Wu, Mingyu Zhang, Jiaxin Li, Sisi Li, Yimin Sheng, Tianxi Tan, Zhenkai Zhang, Jianyi Zhu, Yong-Lu Li
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.08600v2 Announce Type: replace-cross Abstract: World models have recently achieved impressive progress in visual prediction and interactive generation, but extending them to multi-agent environments introduces a fundamental scalability challenge. Existing methods generally assume a fixed ...

📖 Read original article


548. When Do Anchor-Based Pointwise LLM Rerankers Help? Retriever Quality, Statistical Scope, and Anchor Design ​

Author: Utshab Kumar Ghosh, Shubham Chatterjee
Published: 8/18/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.10528v3 Announce Type: replace-cross Abstract: Anchor-based pointwise LLM reranking scores each candidate against a shared reference passage to recover cross-document context at pointwise cost. We study when this actually helps, using GCCP/PAGC as a representative method. Our study is rep...

📖 Read original article


549. When Self-Consistency Backfires: Majority Vote Hurts the Majority of Hard Science Problems for Small LLMs ​

Author: Utkarsh Bahuguna
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.11403v2 Announce Type: replace-cross Abstract: Self-consistency via majority vote reduces per-problem accuracy on most GPQA Diamond problems for small instruction-tuned models: 56.6% of problems for Qwen2.5-7B and 65.7% for Llama-3-8B. The obvious remedy is a verifier-free confidence gate...

📖 Read original article


550. Cross-Corpus Evaluation of Generalizable Vulnerability Detection in IoT Firmware ​

Author: Sadib Hassan Rumman, Md. Shariful Islam, Md Rayhanur Rahman
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.11492v2 Announce Type: replace-cross Abstract: IoT firmware vulnerability detection is constrained by ecosystem heterogeneity, resource-limited platforms, and benchmark quality limitations. Existing datasets are often synthetic or general-purpose and lack human-verified, contamination-scr...

📖 Read original article


551. One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL ​

Author: Simon Yu, Nicholas Tomlin, Marwa Abdulhai, Ximing Lu, Derek Chong, Abe Hou, Dilara Soylu, Sergey Levine, Christopher D. Manning, Weiyan Shi
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.12253v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that this approach systematically fails to generalize, and trace the failure to simulator collaps...

📖 Read original article


552. Excess Separability: Nuisance-Controlled Residual-Stream Probing for Benchmark Contamination Detection ​

Author: Florian Braun
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.12652v2 Announce Type: replace-cross Abstract: Benchmark contamination is diagnosed with n-gram overlap, likelihood-based membership inference, or canary strings, and each needs something usually unavailable: the training corpus, a well-chosen test statistic, or foresight at release. A re...

📖 Read original article


553. MobileMem: Learning from a Year of Mobile Experiences ​

Author: Xinle Deng, Yida Xue, Xiangyuan Ru, Yijun Chen, Buqiang Xu, Mingjun Mao, Xinjie Liu, Haoming Xu, Shuofei Qiao, Mengru Wang, Chen Jiang, Yuchen Eleanor Jiang, Lizhong Wang, Jason Wang, Li Zeng, Haofen Wang, Guilin Qi, Huajun Chen, Ningyu Zhang
Published: 8/18/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA, cs.MM

arXiv:2608.13606v2 Announce Type: replace-cross Abstract: The next generation of AI agents is increasingly moving beyond systems that answer isolated questions toward persistent personal assistants that can understand, remember, and continuously learn from users' experiences. Such assistants require...

📖 Read original article


554. GALA: Generation-Aware Cross-Modal Alignment for Text-to-Time-Series Synthesis ​

Author: Haochen Zhang, Gengwei Zhang, Laura Yao, Nicholas Konz, Tianlong Chen
Published: 8/18/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.13741v2 Announce Type: replace-cross Abstract: Synthesizing time series from natural language is emerging as the most expressive form of controllable time series generation. However, existing text-conditioned generators either take caption embeddings frozen from off-the-shelf text encoder...

📖 Read original article


555. Pairton: Iterative Reconstruction of Short-Lived Particles ​

Author: Andreas Hermansen, Chris Scheulen, Tobias Golling
Published: 8/18/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-ex

arXiv:2608.14278v2 Announce Type: replace-cross Abstract: We present Pairton, an iterative framework for reconstructing short-lived particles in high-energy collision events. By formulating particle reconstruction as a masked prediction process over graph structures, Pairton learns conditional distr...

📖 Read original article