arXiv cs.LG - 2026-08-26 ​
277 items collected.
1. Equivariant Cellular Sheaves for Molecular Electronic Structure: Bridging Sheaf Cohomology and E(3)-Equivariant Hamiltonian Learning ​
Author: Krishna Harish
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph
arXiv:2608.23571v1 Announce Type: new Abstract: Equivariant message-passing networks are the standard model for molecular property and interatomic-potential prediction, and recent work predicts the electronic Hamiltonian itself in an E(3)-equivariant way. Separately, topological deep learning has ex...
2. Data Predictability Shapes Weibull Weight-Scale Growth in Transformer Training ​
Author: Tiexin Ding
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.23573v1 Announce Type: new Abstract: A trained transformer's weight magnitudes can be summarized by a two-parameter Weibull distribution whose shape $k \approx 1.2$ is stable across layers and models, so the scale $\lambda$ carries most training-induced movement. What corpus property sets...
3. From Causal Plausibility to Causal Reliability: Evaluating LLMs as Calibrated Direct Causal-Edge Classifiers ​
Author: Amit Kumar, Elnur Adl Zarabi, Suranjana Trivedy, Zhiqian Chen, Lei Zhang, Kaiqun Fu, Taoran Ji
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ME
arXiv:2608.23660v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to provide prior causal knowledge for structural causal discovery, yet whether their direct-edge judgments and confidence can be trusted remains unclear. We systematically evaluate 12 instruction-tuned...
4. Renormalization Group Flow Matching for Scalable Local Generative Modeling ​
Author: Kanta Masuki, Yuto Ashida
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech
arXiv:2608.23696v1 Announce Type: new Abstract: Despite their remarkable success in modeling complex data, generative models face a fundamental tradeoff. Global approaches can capture full structural coherence but suffer from high computational costs, while local models are efficient but often fail ...
5. Response Renormalization for Critical Deep Equilibrium Models ​
Author: Jose Luis Lima de Jesus Silva
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2608.23725v1 Announce Type: new Abstract: Deep Equilibrium Models (DEQs) compute predictions from a hidden representation unchanged by the model update. Training through this equilibrium uses implicit differentiation and requires solving an adjoint system built from the residual Jacobian. If t...
6. Calibration-Preserving Pruning: Compression as a Reliability Contract ​
Author: Ibne Farabi Shihab, Adria Binte Habib, Anuj Sharma
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.23744v1 Announce Type: new Abstract: Split conformal prediction, not the pruning rule, supplies finite-sample marginal coverage once a pruned model is fixed independently of the conformal calibration split. We study the separate efficiency problem: can pruning preserve score geometry well...
7. Tight Majorizations and Convergence Rates of Nuclear Norm Minimization IRLS ​
Author: Christian K"ummerle, Tomas Masak, Dominik St"oger
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, math.OC
arXiv:2608.23765v1 Announce Type: new Abstract: Iteratively reweighted least squares (IRLS) methods constitute a natural approach to nuclear norm minimization, but their convergence rates and the role of the weight operator have remained poorly understood. This paper establishes sharp convergence ra...
8. Disentangled Skill Representations for Predictive Human Modeling ​
Author: Mariah Schrum, Deepak Gopinath, Srijan Srivatsa, Guy Rosman, Tiffany Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.23776v1 Announce Type: new Abstract: Understanding human skill is important for AI systems that collaborate with, coach, or assist people. Unlike typical latent variable estimation problems which rely on single observations, skill is a persistent, compositional, and behaviorally grounded ...
9. GAP-Prompt: Gated Adaptive Prompting for Efficient Continual Learning ​
Author: Trung-Anh Dang, Duy-Cuong Bui, Ngoc-Son Vu, Christel Vrain, Vincent Nguyen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23782v1 Announce Type: new Abstract: Continual learning faces the persistent challenge of catastrophic forgetting, where sequential task updates degrade previously acquired knowledge. While prompt-based methods integrated with pre-trained models offer a compelling solution by freezing the...
10. Mixture of Channel Experts: Static Sparse Supports with Input-Adaptive Mixing for Pointwise Projections ​
Author: Elian Iluk, Gil Ben-Artzi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.23794v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) scales language models by routing each input through a small set of independently parameterized experts. We show that copying this design into convolutional networks fails for a structural reason: parallel convolutional experts...
11. A Theory of Speciation in Generative Diffusion Models on Compact Riemannian Manifolds ​
Author: Alessio Marta, Paola Causin
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23798v1 Announce Type: new Abstract: Speciation in generative diffusion models denotes the emergence of distinct stable branches during denoising, through which initially undifferentiated trajectories progressively commit to different data classes. In this work we develop an intrinsic the...
12. Discovering Cross-Language Reasoning Invariance in LLMs with Geometry-Invariant Sparse Autoencoders ​
Author: Igor Bogdanov, Changcheng Huang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.23809v1 Announce Type: new Abstract: Multilingual language models can solve the same mathematical problem in different languages, but it remains unclear whether they rely on shared features or on language-specific computations that only produce similar outputs. We study this question in f...
13. Learning to Grade Efficiently: A Bandit-Driven Prompt-Selection Framework for Low-Cost LLM Essay Scoring ​
Author: Olga Manakina, Igor Bogdanov
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.23814v1 Announce Type: new Abstract: Large Language Models (LLMs) demonstrate strong capabilities in automated essay scoring (AES), but contemporary approaches typically employ fixed prompt selection, failing to address operational cost concerns and evolving optimal configurations. We pro...
14. AQLoRA: A Zero-Search Recipe for Fast Quantized LoRA Fine-Tuning ​
Author: Md Romyull Islam
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23816v1 Announce Type: new Abstract: Quantized fine-tuning (QLoRA) saves memory but not time. It dequantizes every 4-bit weight on the fly, so it trains more slowly than fp16 LoRA. We present AQLoRA (Adaptive-Quantization LoRA), a recipe that buys part of that time back. One CPU pass over...
15. Generating Intervention Hypotheses using Explainable Explanations on Graphs: G2I, a Two-Stage Greedy Framework ​
Author: Mulin Tian, Ajitesh Srivastava
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2608.23835v1 Announce Type: new Abstract: Real-world decision-making in public health and social science can greatly benefit from predictive models, yet translating predictions into effective interventions requires explaining the model behavior. While Graph Neural Networks (GNNs) are well-suit...
16. PuzzleKV: Page-Wise Low-Rank Decomposition for KV Cache Compression ​
Author: Zizhong Wang, Jieying Wang, Zhao Zhang, Jiajia Li
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23843v1 Announce Type: new Abstract: Long-context inference in large language models (LLMs) is increasingly limited by the memory required for the key-value (KV) cache. KV cache compression addresses this problem by reducing the storage cost of previous tokens. Among existing approaches, ...
17. FlowNeg: GFlowNet-Guided Diverse Hard Negative Sampling for Knowledge Graph Embedding ​
Author: Ibne Farabi Shihab, Naoshin Anzum Hridi, Joyanta Jyoti Mondal
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23849v1 Announce Type: new Abstract: Negative sampling determines whether a knowledge graph embedding (KGE) model learns from informative counterexamples or wastes updates on implausible corruptions. Uniform negatives are diverse but easy, whereas hard-negative miners concentrate on few e...
18. UHI-Bench: Benchmarking Dual-Source Urban Heat Island Modeling Across Cities in Diverse Climate Regimes ​
Author: Wanyun Ling, Chenxi Liu, Yi Xie, Aopu Xu, Zhuoqi Zeng, Ziyue Li
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23857v1 Announce Type: new Abstract: Urban heat islands (UHIs) are intensifying under climate change, exacerbating thermal exposure risks. Their two primary observations, land surface temperature UHI (LST-UHI) and near-surface air temperature UHI (AirT-UHI), capture physically distinct as...
19. Revelation Control ​
Author: Qinyou Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.23860v1 Announce Type: new Abstract: Revelation Control is the problem of choosing priced interventions that reveal hidden state only insofar as the revealed distinctions can change a consequential decision, while accounting separately for any useful progress created by the intervention i...
20. Every Layer Counts: An Exponential $L_2$ Depth Hierarchy for ReLU Networks ​
Author: Itay Safran
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23877v1 Announce Type: new Abstract: We prove a depth hierarchy for ReLU neural networks in which every additional ReLU layer can save exponentially many neurons. For every $\ell\geq 3$, a globally $[0,1]$-valued, $1$-Lipschitz function is realized by a depth-$\ell$ network of width $\mat...
21. Partial Optimal Transport on the Circle for All Transported Masses in O(N log N) ​
Author: Soheil Kolouri
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.DS
arXiv:2608.23910v1 Announce Type: new Abstract: Partial optimal transport compares two measures while leaving part of the mass unmatched, which is what makes it robust to outliers, occlusion, and clutter. The quantity of interest is usually the whole profile - the optimal cost at every transported c...
22. The Loss Floor of Denoising Score Matching: Fisher Geometry from Schr\"odinger Bridges ​
Author: Avinash Raju, Kai Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech
arXiv:2608.23916v1 Announce Type: new Abstract: Denoising score matching trains diffusion models by regressing onto a conditional score, although generation ultimately requires the marginal score. The two objectives share the same population minimizer, but the conditional target remains random at fi...
23. GATNextHop: A GAT for Shortest Path Routing with Cross-Topology Generalization ​
Author: Chia-Hong Chou, Katerina Potika
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23917v1 Announce Type: new Abstract: Common shortest-path algorithms, such as Dijkstra's (SPF), that OSPF uses, provide exact routing solutions but must be recomputed for each network topology, limiting scalability in dynamic or large-scale networks. This paper proposes the GATNextHop mod...
24. MnemoDyn: Learning Resting State Dynamics from 40K FMRI sequences ​
Author: Sourav Pal, Viet Luong, Hoseok Lee, Tingting Dan, Guorong Wu, Richard Davidson, Won Hwa Kim, Vikas Singh
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23936v1 Announce Type: new Abstract: We present a dynamical-systems based model for resting-state functional magnetic resonance imaging (rs-fMRI), trained on a dataset of roughly 40K rs-fMRI sequences covering a wide variety of public and available-by-permission datasets. While most exist...
25. CoDrift: Compositional Drifting for Offline Reinforcement Learning ​
Author: Xiewei Ni, Ruofeng Mei, Xiangyu Xu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2608.23939v1 Announce Type: new Abstract: Offline reinforcement learning is intrinsically multi-objective: a policy must remain compatible with the behavioral support of a fixed dataset while preferentially selecting high-value actions. We recast these objectives in a common form by viewing ea...
26. Low-Latency Activation-Regularized Sparse Neural Operators with Distillation Assistance Towards Real-Time Edge-Deployable Virtual Sensing ​
Author: William Howes, Farid Ahmed, Syed Bahauddin Alam
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.23987v1 Announce Type: new Abstract: Virtual sensing enables digital twins and safety-critical systems to reconstruct and forecast spatial-temporal physics in real time. However, conventional computational and data-driven methods often face challenges in generalization, latency, and energ...
27. Revenge of Monosemanticity: Specialized Neurons Improve Data Efficiency in MLPs ​
Author: Amirhesam Abedsoltan, Enric Boix-Adsera, Fivos Kalogiannis, Mikhail Belkin
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.24007v1 Announce Type: new Abstract: Understanding how neural networks learn and organize features is central to understanding their behavior. Much existing theory of feature learning has focused on the emergence of a global low-dimensional predictive geometry. We show that this picture i...
28. ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning ​
Author: Juntao Fang, Shifeng Xie, Ruichu Cai, Shengji Zheng, Zijian Li, Keli Zhang, Lujia Pan, Themis Palpanas, Zhifeng Hao
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.24033v1 Announce Type: new Abstract: Time series classification underpins applications in healthcare, sensing, and industrial monitoring. Although time series foundation models support forecasting and transferable representation learning, classification still typically requires fitting a ...
29. PinSieve: Production Selective VLM Serving and a Governed Memory Flywheel for Enterprise Content-Quality Triage ​
Author: Chuqing Gao, Yuanfang Song, Jonathan Zhang, Yifan Wu, Vishwakarma Singh, Qinglong Zeng, Andrey Gusev
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24040v1 Announce Type: new Abstract: Enterprise AI agents in production often need to be bounded, stateful, observable, and governable rather than fully autonomous. We present PinSieve, a production case study in a large-scale content-quality pipeline. Its deployed component is a selectiv...
30. XP-JEPA: Cross-Predictive Physics Grounding for Forecastable Latent Dynamics ​
Author: Kehan Wen, Ziming Li, Siyuan Luo, Fan Shi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24044v1 Announce Type: new Abstract: Latent world models plan by predicting how candidate actions transform learned representations. In self-predictive models, however, the encoder and predictor are optimized jointly and can co-adapt to latent transitions that are easy to predict but only...
31. Physics-Integrated Operator Learning via Gaussian Splatting Representations ​
Author: Jihao Zhang, Junyi Guo, Jian-Xun Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24049v1 Announce Type: new Abstract: Neural operators provide efficient surrogates for spatiotemporal PDE systems, but purely data-driven formulations often accumulate substantial errors during long-horizon autoregressive prediction and may fail to exploit available governing-equation str...
32. ALPHABET: A Laplace-Pole History Aggregator with Banked Exponential Transport ​
Author: Daehwa Ko, JaeHyeon Kim, Oh Seong Kwon, Jay Hoon Jung
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24051v1 Announce Type: new Abstract: Can a sequence model remain competitive with only a few thousand parameters and an explicitly auditable prediction interface? We introduce ALPHABET, a compact linear-time model that compresses temporal history into stable complex pole modes: a direct b...
33. PhysicsBench: A Unified Leaderboard for Generative and Predictive Models in Engineering Design and Simulation ​
Author: Sang Won Lee, Hyogu Jeong, Namwoo Kang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2608.24056v1 Announce Type: new Abstract: Generative and predictive artificial intelligence models are increasingly used to generate geometry and to predict physical fields and scalar quantities in engineering design and simulation. Yet these models are typically evaluated in isolation, on aca...
34. Mechanistic Circuit Identification for Controllable Data Generation ​
Author: Nakyung Lee, Sangwoo Hong, Jungwoo Lee
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.24065v1 Announce Type: new Abstract: While recent advances in data synthesis aim to curate high-quality datasets, most generation pipelines still rely on heuristic prompt-based control. This black-box paradigm provides limited insight into how individual samples interact with a model's un...
35. A Feature-Major Codebook for Memory-Efficient Sparse-Binary Self-Organizing Maps: Scaling a MEDLINE Atlas to 1.05 Million Neurons on a Single Consumer GPU ​
Author: Andrew James Amos
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2608.24067v1 Announce Type: new Abstract: A self-organising map turns a large corpus into a browsable two-dimensional atlas, but building one at MEDLINE scale has been impractical: the best-matching-unit (BMU) search that dominates training is bound by the bandwidth needed to read the codebook...
36. Knowing When to Ask for Help: Bayesian Self-Escalation in Hierarchical LLM Agents ​
Author: Nadeem Shaikh
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.24087v1 Announce Type: new Abstract: Current LLM agent systems decide delegation before reasoning begins (a router picks a model) or after a response is complete (a verifier scores it and may retry). We study a third regime: an agent that recognises, during its own reasoning, that it is u...
37. The Sharp Tail of Uniform Stability ​
Author: Pahan Dewasurendra
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24098v1 Announce Type: new Abstract: Uniform stability controls how much one training example can change the loss at any test point. A new logarithmic-free upper bound shows that a $\gamma$-uniformly stable algorithm with loss in $[0,L]$ has generalization gap at most $O \left(\gamma\log(...
38. Structured Frequency-Domain Evidence for LLM-Based Time-Series Anomaly Detection ​
Author: Jungwook Seo, Sangwon Son, Minjeong Kim, Seungmin Han, Seojin Yoo, Sungyong Baik
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24113v1 Announce Type: new Abstract: Time-series anomalies can appear not only as pointwise deviations but also as changes in recurring temporal structure, such as shifted periodicity or localized oscillatory fluctuations. However, existing LLM-based time-series anomaly detection methods ...
39. A mesh-free multiresolution deep energy method with phase-field modeling of brittle fracture ​
Author: Han Zhang, Mehrisadat Makki Alamdari, Babak Shahbodagh, Mohammad Vahab, Cosmin Anitescu, Timon Rabczuk, Elena Atroshchenko
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.24126v1 Announce Type: new Abstract: Phase-field modeling of brittle fracture removes the need to track cracks explicitly by recasting their evolution as the minimization of an energy functional. In return it requires a discretization dense enough to resolve a localization band whose widt...
40. From Gradient-Boosted Trees to Deep Recommenders: Practical Lessons from Migrating a Production Customer Support Recommender ​
Author: Sonia Sharma, Jeyendran Balakrishnan, Shreya Rajpal, Swapnil Parekh, Nagaraj Janardhana, Andrew Mattarella-Micke
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24132v1 Announce Type: new Abstract: Product catalogs in fast-moving service businesses are shifting from static, independently priced SKUs toward dynamically bundled, discount-coupled offerings--a shift that strains the tree-based classifiers traditionally preferred for sparse and highly...
41. Steering Recurrent Reasoners at Inference Time with Readout Feedback ​
Author: Shunsuke Kamiya, Masanori Koyama, Seongcheol Jeong, Fumiya Uchiyama, Kenji Kubo, Kohei Hayashi, Masahiro Suzuki, Yutaka Matsuo
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24136v1 Announce Type: new Abstract: Recurrent models, which repeatedly update latent states with shared computation blocks, have emerged as powerful architectures for solving complex reasoning tasks. Existing inference-time methods scale computation by running more steps or sampling more...
42. Robust Data-Collection Policy Learning for Low-Variance Online Policy Evaluation ​
Author: Claire Chen, Shuze Daniel Liu, Licheng Luo, Rohan Chandra, Nan Jiang, Shangtong Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.24146v1 Announce Type: new Abstract: In reinforcement learning policy evaluation, classic on-policy methods often suffer from high variance when estimating policy performance. To mitigate this issue, behavior policy search has been proposed to learn data-collecting policies tailored to re...
43. From Relaxed Indexability to Exact Indexability: A $t$-Step Approach for Partially Observable Restless Bandits ​
Author: Qizhen Jia, Keqin Liu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2608.24167v1 Announce Type: new Abstract: Whittle index policies offer a scalable method for restless multi-armed bandits, but under partial observability even determining the indifference subsidy at a single belief requires solving an infinite-horizon belief-state problem with no closed-form ...
44. PRQ-KMeans: Projection Residual Quantization for Semantic ID Tokenization ​
Author: Yunxiao Luo, Siyuan Wang, Ben Chen, Chenyi Lei
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24207v1 Announce Type: new Abstract: Semantic identifiers (SIDs) represent entities as hierarchical token sequences for generative retrieval and recommendation. Residual-quantization tokenizers construct these sequences by selecting a codeword at each level and passing a residual to the n...
45. A Data-dependent Early Stopping Rule using Rademacher Complexity with L1-norm ​
Author: Duy Hoang, Bastien Berret, Olivier Bruneau, Laurent Fribourg
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24210v1 Announce Type: new Abstract: Training neural networks requires balancing the trade-off between fitting the training data and achieving robust performance on unseen inputs. This ability, commonly referred to as generalizability, is determined by the gap between the empirical risk o...
46. Contrastive Branch Policy Optimization ​
Author: Ying Wang, Changlin Qiu, Bang Lin, Linbo Jin, Wen Jiang, Zhe Sun, Jingli Yang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24300v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) enables language models to learn multi-turn interaction with external tools, yet its sparse outcome rewards provide no signal for identifying which intermediate decisions are responsible for success...
47. Causal Analysis for Time Series Foundation Models ​
Author: Mathis Jander, Wouter van Heeswijk, Martijn Mes
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24303v1 Announce Type: new Abstract: Transitioning from bespoke time series models towards time series foundation models changes the relationship of model and application from one-to-one to one-to-many. This shift introduces concentration risk as many, potentially high-risk, forecasting a...
48. A Structural FHMM for Interpretable Disease Trajectories in T2DM ​
Author: Alessandro Mari, Ekaterina Krymova, Guillaume Obozinski, Maria Luisa Marques de Sa Faquetti, Adrian Martinez de la Torre, Andrea Burden
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24328v1 Announce Type: new Abstract: In this work, we propose a structural variant of the Factorial Hidden Markov Model (FHMM) for the analysis of disease trajectories in patients with Type 2 diabetes mellitus (T2DM). The model represents a patient's latent health state as a combination o...
49. When Does Self-Supervised Pretraining Help Tabular Models? A Study of Label Scarcity and Missing Data ​
Author: Sahand Mazrouei
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24381v1 Announce Type: new Abstract: Self-supervised learning (SSL) has emerged as a promising approach for tabular data, yet its efficacy under extreme label scarcity and test-time missingness remains under-explored. In this paper, we evaluate a mask-and-recover SSL pretraining objective...
50. Equivariant Covariance Tensors: Guaranteed SPD Uncertainty for Tensor-Valued Geometric Learning ​
Author: Ruihan Liu, Yu Ji, Jianbo Yu, Shifu Yan, Qingchao Jiang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24386v1 Announce Type: new Abstract: Tensor-valued prediction is fundamental to geometric deep learning, yet uncertainty quantification (UQ) for such outputs remains an open challenge. While E(3)-equivariant neural networks excel at point estimates, they lack rigorous confidence measures....
51. Joint Distribution Alignment for Universal Domain Adaptation ​
Author: Shizhe Li, Hongshan Pu, Mengying Xie, Yi Xiang, Xiaowei Yang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.24429v1 Announce Type: new Abstract: Unsupervised domain adaptation (UDA) has been widely concerned in the fields of machine learning, pattern recognition, and computer vision. Traditional UDA learning usually assumes that the label spaces of the source and target domains are exactly the ...
52. Evaluating Deep Multivariate Imputation Models on Wearable Device Data ​
Author: Skye Goodman, Roussel Desmond Nzoyem, Leandro Junges, Peter Kissack, Yasser Qureshi, Amberly Brigden, Jeff Clark, Nawid Keshtmand
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.QM
arXiv:2608.24436v1 Announce Type: new Abstract: Wearable device data enables continuous health monitoring, but suffers from structured missingness: features sharing a physical sensor drop out together. Deep imputation methods such as BRITS and SAITS have seen limited evaluation on multimodal physiol...
53. WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation ​
Author: Zihao Wu, Hongyao Tang, Yi Ma, Huizhong Song, Pengyi Li, Yifu Yuan, Fei Ni, Jinyi Liu, Wei Wei, Jianrong Wang, Yan Zheng, Jianye Hao
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24479v1 Announce Type: new Abstract: Massively parallel simulation changes the data regime in which off-policy reinforcement learning (RL) is trained, challenging stabilizers designed for data-limited replay. Through controlled experiments across eight benchmark families, we show that the...
54. Beyond Static Interpretability: Anticipating Post-SFT Mechanisms from Pre-SFT Parameters for Better Tuning ​
Author: Hang Chen, Jiaying Zhu, Wenya Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.24482v1 Announce Type: new Abstract: Mechanistic Localization bridges mechanistic interpretability and post-training optimization by isolating critical parameters via interpretative approaches and then guiding parameter-efficient Supervised Fine-Tuning (SFT) in a ``locating-then-tuning'' ...
55. Where Entropy Is Measured Matters: Policy Geometry in Bounded Continuous-Control PPO ​
Author: Yiyang He, Zhichun Zhou, Ziwei Wang, Tao Xue, Haolin Fei
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24488v1 Announce Type: new Abstract: Many continuous-control policies are optimized as unbounded Gaussians and then mapped into bounded actions. We show that where entropy is measured changes the policy geometry learned by proximal policy optimization (PPO). In an 80-muscle MyoLeg task, a...
56. When Do Supervised UQ Ensembles Improve LLM Hallucination Detection? A Robustness Study ​
Author: Mohit Singh Chauhan, Vipin Gyanchandani, Dylan Bouchard
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.24492v1 Announce Type: new Abstract: Uncertainty quantification (UQ) methods are widely used for hallucination detection in large language models (LLMs) in closed-book settings where ground-truth evidence is unavailable at inference time. Prior work has proposed combining UQ signals via l...
57. It depends: Incorporating correlations for joint aleatoric and epistemic uncertainties of high-dimensional output spaces ​
Author: Leonhard F. Feiner, Manuel Nickel, Martin Menten, Laurin Lux, Rickmer Braren, Daniel Rueckert, Georgios Kaissis, Raphael Rehms, Johannes Paetzold
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.24518v1 Announce Type: new Abstract: Uncertainty Quantification (UQ) plays a vital role in enhancing the reliability of deep learning model predictions, especially in scenarios with high-dimensional output spaces. This paper addresses the dual nature of uncertainty -- aleatoric and episte...
58. From Numerical Simulators of PDEs to Neural Emulators and Back ​
Author: Felix Koehler
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24547v1 Announce Type: new Abstract: Simulation is central to modern engineering and science, but the cost of numerical solvers for partial differential equations (PDEs) remains a bottleneck whenever fast or many-query evaluations are required. Neural emulators trained on solver-generated...
59. Persistent Cross Entropy ​
Author: Sijin Yeom, Jae-Hun Jung
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2608.24549v1 Announce Type: new Abstract: Persistent entropy is the Shannon entropy of a persistence-based probability measure defined on a persistence diagram. However, its cross-entropy version is not naturally defined because two persistence diagrams generally have different event spaces. T...
60. FraudBench: Protocol-Sensitive Benchmarking of Adversarial Robustness for Financial Risk Assessment ​
Author: Xitong Zeng, Zhaoge Bi, Yitian Yang, Huaming Chen, Quan Z. Sheng
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24551v1 Announce Type: new Abstract: Machine learning models are widely used in financial fraud and credit-risk detection, yet their adversarial robustness remains difficult to evaluate because financial tabular data involve domain-specific constraints, severe class imbalance, and asymmet...
61. SeisMamba: Low-Latency Single-Station Seismic Magnitude Estimation for Spatially Distributed Earthquake Early Warning ​
Author: Quenton Yeo, Zhaoge Bi, Linghan Huang, Luke Stephen Higgins, Flora Salim, Huaming Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24561v1 Announce Type: new Abstract: Rapid earthquake magnitude estimation is central to earthquake early warning, yet many operational systems depend on dense regional seismic networks and region-specific calibration. This creates a spatial coverage barrier for high-risk areas with spars...
62. Across the Loss Landscape with Progressive Growth ​
Author: Paul Caillon, Christophe Cerisara, Alexandre Allauzen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24568v1 Announce Type: new Abstract: Deep neural networks generalize well despite their highly nonconvex, overparameterized loss landscapes, a phenomenon often associated with the geometry of the minima found by stochastic optimization. We study how incremental grow-and-optimize strategie...
63. IAPO: Influence-Aware Policy Optimization for Credit Assignment in Multi-Turn Service Agents ​
Author: Bo Ren, Yirong Mao, Yi Yang, Wenhui Que
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24588v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly solve long-horizon tasks through multi-turn interactions with users and external tools. In these settings, relevant task information often unfolds over time rather than being fully specified at the initial...
64. Delayed Optimizer-State Transport Shapes Short-Horizon Training Decisions ​
Author: Jinhui Guo
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2608.24593v1 Announce Type: new Abstract: Adaptive optimizers retain gradient history in moment variables, allowing a local change in loss weighting to alter later updates. We examine whether this delayed transport is large enough to change prospective short-horizon decisions. On committed fut...
65. Taming foundation model with invariance-oriented pre-training for broad-spectrum EEG analysis across signal-level, brain-state, and brain-health tasks ​
Author: Yulong Dou, Han Wu, Guo Chen, Fangmao Ju, Zhiming Cui, Dinggang Shen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24597v1 Announce Type: new Abstract: Electroencephalography (EEG) is a widely used window into human brain function, but most EEG models remain tied to a one-dataset-one-model supervised paradigm. Recent EEG foundation models offer a route toward reusable representations, but most remain ...
66. Conditional GraphGANFed: Optimizing Graph-Structured Molecule Generation in Federated Generative Adversarial Networks ​
Author: Daniel Manu, Abee Alazzwi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2608.24610v1 Announce Type: new Abstract: Generative adversarial networks (GANs) have garnered considerable attention in molecular discovery for their ability to generate novel and high-quality molecules. To efficiently train a GAN model while preserving data privacy, GraphGANFed has been prop...
67. Bandit Submodular Maximization under Matroid Constraints: Learning Compressed Exchange Policy ​
Author: Zongqi Wan, Zhijie Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24627v1 Announce Type: new Abstract: We study adversarial bandit maximization of monotone submodular functions under a matroid constraint. For a rank-$k$ matroid on $n$ elements, we give a randomized oracle-polynomial algorithm that makes one feasible value query per round and has expecte...
68. Data Leakage Inflates Generalizability of Power Outage Prediction Models ​
Author: Yamil Essus, Ranga Raju Vatsavai, Benjamin Rachunok
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24665v1 Announce Type: new Abstract: Power outage prediction models are increasingly used in assessments of climate-driven infrastructure risk, yet current evaluation practices obscure whether these models generalize to the novel conditions such applications require. We identify three com...
69. A Multimodal Foundation Model for Longitudinal Patient Representation and Scalable Insight Generation in Oncology ​
Author: Eugene Vorontsov, Yi Kan Wang, Alican Bozkurt, Adam Casson, Ludmila Tydlitatova, Michal Zelechowski, Ezra E. W. Cohen, Jyoti D. Patel, Max Banaszak, Caitlin McWilliams, Shane Colley, Kate Sasser, Ryan Fukushima, Eric Lefkofsky, Razik Yousfi, Siqi Liu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24688v1 Announce Type: new Abstract: Precision oncology necessitates a longitudinal model of patient state that captures cancer evolution and treatment over time, integrating multimodal observations. We introduce the oFM, a foundation model developed on a real-world oncology cohort of 1.6...
70. On-policy Distillation with Verifiable Reward ​
Author: Wenze Lin, Jiale Zhao, Xitai Jiang, Songde Rao, Yining Li, Shenzhi Wang, Bingxiang He, Gao Huang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24696v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) and on-policy distillation (OPD) have become two widely adopted paradigms for post-training large language models. However, RLVR suffers from sparse task-level feedback, while OPD provides dense tok...
71. Single State Update Predictive Coding training for Time Series Forecasting and Anomaly Detection ​
Author: Matteo Cardoni, Sam Leroux
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2608.24697v1 Announce Type: new Abstract: Predictive Coding (PC) is a neural learning paradigm that enables parallelizable neural network layer updates. However, the main bottleneck of PC Networks (PCN) is the sequential backwards error propagation. To tackle this, we introduce a training tech...
72. Parameter-Level Attribution of Symmetry in Trained Networks Though Parameter-Wise Functional Sensitivity ​
Author: Alan Muriithi, Vedanta Thapar, Torben Berndt
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24700v1 Announce Type: new Abstract: When a network has learned a function with a known symmetry, can that symmetry be moved through the parametrisation---is there a motion in parameter space realising the group action in function space? We formulate this as a lifting problem for the real...
73. Constrained Hyperparameter Optimization for Streaming Data ​
Author: Bruno Veloso, Jo~ao Gama
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24712v1 Announce Type: new Abstract: Optimization of hyperparameters is a critical factor to obtain optimal model performance. While existing research has predominantly concentrated on batch-learning scenarios, addressing the complexities inherent in data streams presents a challenge. The...
74. Enhancing Bayesian Optimization and Active Learning Through Kernel Diversity ​
Author: Heng Zhang, Haotian Xiang, Qin Lu, Konstantinos D. Polyzos, Tara Javidi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24721v1 Announce Type: new Abstract: Hyperparameter selection remains a key challenge in Bayesian optimization (BO) and Bayesian active learning (AL), as model misspecification can lead to suboptimal performance, while more accurate fully Bayesian treatments typically rely on computationa...
75. Parameter-Efficient Self-Supervised Adaptation for EEG-FM under Fixed Computational Budgets ​
Author: Meghal Dani, Stefanie Liebe
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24727v1 Announce Type: new Abstract: EEG foundation models pretrained via self-supervised learning promise transferable representations, but their generalization remains limited, especially across diverse clinical datasets. Full fine-tuning is impractical for resource-constrained clinical...
76. Optimal Alternating Regret for Online Learning and Games ​
Author: Yixin Tao, Weiqiang Zheng
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.GT, stat.ML
arXiv:2608.24731v1 Announce Type: new Abstract: We settle the minimax-optimal alternating regret, a regret notion motivated by alternating learning dynamics in games, for both online linear optimization (OLO) and online convex optimization (OCO). For OLO over the probability simplex $\Delta_d$, we g...
77. $(\text{DNN})^2$: Doubly Non-Negative Relaxations for Deep Neural Networks ​
Author: Hanna Jiamei Zhang, Alan Papalia, Michael Everett, David M. Rosen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.RO, cs.SY, eess.SY
arXiv:2608.24743v1 Announce Type: new Abstract: Existing linear program (LP) and semidefinite program (SDP) relaxations for rectified linear unit (ReLU) neural network (NN) verification yield overly-conservative safety guarantees due to significant relaxation gaps. While the completely positive prog...
78. Beyond Uniform Local Isometry and Topology: FactoMap for Disentangled Representations ​
Author: Sohini Gupta, Bahareh Tolooshams
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.24762v1 Announce Type: new Abstract: Many disentanglement methods represent generative factors using Euclidean product coordinates, although the underlying factor spaces may wrap, collapse, or have position-dependent geometry. We introduce factor-space structure, combining factor domains,...
79. LION: A Clifford Neural Paradigm for Multimodal-Attributed Graph Learning ​
Author: Xunkai Li, Zekai Chen, Zhengyu Wu, Henan Sun, Daohan Su, Guang Zeng, Hongchao Qin, Rong-Hua Li, Guoren Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24795v1 Announce Type: new Abstract: Recently, the rapid advancement of multimodal domains has driven a data-centric paradigm shift in graph ML, transitioning from text-attributed to multimodal-attributed graphs. This advancement significantly enhances data representation and expands the ...
80. MDTE: Minority-Aware Diffusion over Temporal Edge Events for Imbalanced Node Classification ​
Author: Zhou Zelong, Zhang Tianming, Yang Zhengyi, Tang Yifu, Hou Chenyu, Cao Bin, Fan Jing
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24812v1 Announce Type: new Abstract: Class-imbalanced node classification on temporal graphs is challenging because majority-dominated temporal propagation progressively assimilates minority representations, while conventional node and neighborhood information provides insufficient discri...
81. Effective Learning Rate Governs Loss Dynamics in Language Model Pretraining ​
Author: Zihan Liu, Ruiheng Zheng, Shaobo Zhang, Changxin Tian, Kunlong Chen, Zhiqiang Zhang, Lei Wu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24814v1 Announce Type: new Abstract: We uncover ELR collapse in language model pretraining: learning rate (LR) and parameter norm govern loss dynamics primarily through their ratio, the effective learning rate (ELR). When ELR is matched across runs, their loss trajectories collapse throug...
82. A Geometric Theory of Robust Fairness Audits ​
Author: Binita Maity
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24818v1 Announce Type: new Abstract: Neighborhood-based fairness audits evaluate individual fairness by comparing predictions among similar individuals in feature space. Despite their widespread use, little is known about the robustness of the auditing procedure itself. Because these audi...
83. BioKERN: Biological Kernel Regularization for Histology-to-Transcriptomics Neighborhood Retrieval ​
Author: Seungik Cho, Betul Orcan-Ekmekci
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM
arXiv:2608.24823v1 Announce Type: new Abstract: Spatially resolved biology requires representations that preserve biological neighborhood structure rather than only exact cross-modal correspondences. Existing histology--transcriptomics objectives can emphasize instance-level matching even when non-p...
84. Bellman Calibration for Marginalized Importance Weighting in Offline Reinforcement Learning ​
Author: Lars van der Laan, Nathan Kallus
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.24858v1 Announce Type: new Abstract: Marginalized importance weighting evaluates a target policy by reweighting offline state-action samples with its discounted occupancy ratio, characterized by an adjoint Bellman equation. Existing minimax, primal-dual, and fitted fixed-point estimators ...
85. Improving Cross-Problem Vehicle Routing with Locally Augmented Preferences and Representation Disentanglement ​
Author: Arthur Corr^ea, Paulo Nascimento, Samuel Moniz
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.24859v1 Announce Type: new Abstract: Multi-task vehicle routing problem (VRP) solvers seek to handle multiple VRP variants within a single unified model, avoiding the need to train a separate model for every variant. In spite of recent progress, current approaches remain limited on two fr...
86. Symbolic Classification-Enabled LHC Limits Online BSM Global Fits ​
Author: Shehu AbdusSalam
Published: 8/26/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, cs.SC, hep-ex, hep-th
arXiv:2605.22330v1 Announce Type: cross Abstract: Global fits of Beyond the Standard Model (BSM) physics often involve a two-way interplay between theory and experiment. Theoretical models provide guidance for experimental searches, while experimental results, in turn, constrain theoretical framewor...
87. Finite-Sample Metric Non-Collapse for Geometrically Supervised Latent World Models in Control ​
Author: Alain Bensoussan, Minh-Nhat Phung, Minh-Binh Tran
Published: 8/26/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2608.07265v2 Announce Type: cross Abstract: We establish a finite-sample learning-to-control theory for geometrically supervised latent models of nonlinear deterministic systems. Geometric supervision is used only during training: simulator state, proprioception, or state estimates with indepe...
88. DiD It in 87 Minutes: A Label-Free Softmax-to-Linear Adaptation of Vision Transformers for Object Detection ​
Author: Huaiyuan Qin, Gabriel James Goenawan, Zihang Lin, Muli Yang, Hongyuan Zhu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.22368v1 Announce Type: cross Abstract: While linear attention is a compelling mechanism for high-resolution object detection due to its reduced cost for global token mixing, converting the Softmax-attention ViT backbone of a trained detector into a linear-attention one is not a trivial dr...
89. InfoDPP-PAC: Principled Patch Selection for Whole Slide Image Analysis ​
Author: Prateek Mittal, Ayush Srivastava, Joohi Chauhan
Published: 8/26/2026, 4:00:00 AM
Categories: q-bio.QM, cs.CV, cs.IT, cs.LG, math.IT
arXiv:2608.23574v1 Announce Type: cross Abstract: Each WSI slide contains thousands of candidate tissue patches, while supervision is usually available only at slide level. Existing bag-construction strategies like Uniform extraction and handcrafted heuristics do not control redundancy while attenti...
90. Transformer Accelerator (TFA): A Macro-Op INT8 Hardware Chip for Transformer Inference and Machine Translation ​
Author: Shashank
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AR, cs.CL, cs.LG
arXiv:2608.23582v1 Announce Type: cross Abstract: We present the Transformer Accelerator (TFA), a synthesizable, parameterizable INT8 memory-to-memory engine for transformer inference. One time-multiplexed datapath handles prompt processing and autoregressive generation. TFA implements matrix multip...
91. StateTune: Transforming LLM-Assisted EDA Flow Tuning into a Stateful, Closed-Loop Process ​
Author: Kunlong Li, Shangshang Yao, Su Zheng, Lingli Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, cs.MA
arXiv:2608.23601v1 Announce Type: cross Abstract: EDA flow parameter tuning is critical for quality-of-results~(QoR), yet the parameter space is large, tightly coupled, and full evaluations are prohibitively expensive. Prior LLM-assisted tuners mainly use the LLM as an external proposer with transie...
92. When May an Agent Stop? Evidence-Carrying Termination for Tool-Using LLMs ​
Author: Jason Liu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG
arXiv:2608.23623v1 Announce Type: cross Abstract: Tool-using agents must decide when to stop. Existing systems already gate terminal success, certify execution traces, or enforce runtime polici es, but do not test this particular receipt-, scope-, and closed-replay design at the COMPLETE boundary ac...
93. The Blending Ratio Is Not Where the Performance Is: Diagnosing Prototype Blending for Few-Shot Adaptation of Vision-Language Models ​
Author: Liangzhi Li, Bowen Wang, Yiming Qian, Thorsten Neumann, Xia Xie, Guangshun Li
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.23634v1 Announce Type: cross Abstract: Many few-shot adaptation methods for vision-language models classify with a convex combination of the zero-shot text prototype and the mean of the K labelled image features, with a single blending ratio routinely tuned on held-out labels, often on th...
94. Replicable Conformal Prediction ​
Author: Marios Papamichalis, Regina Ruane, Theofanis Papamichalis
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.23638v1 Announce Type: cross Abstract: Two analysts who calibrate the same predictive model on independent samples will deploy different prediction sets every time, because the calibration threshold inherits the randomness of the data. Wherever deployments must be audited, cached, or appr...
95. Contextual Embedding Evidence for Main--Light Verb Distinctions in Urdu ​
Author: Farah Adeeba, Miriam Butt
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.23645v1 Announce Type: cross Abstract: Urdu light verbs contribute schematic event-structural meaning while remaining lexically related to corresponding main verbs. This study tests representational predictions derived from Butt's analysis using contextual embeddings from UrduBERT, Dunbaa...
96. MolEmb: Multimodal Large Language Models Can Be Strong Molecular Embedding Models ​
Author: Xinjian Zhao, Xiangru Jian, Yaoyao Xu, Xiaozhuang Song, Wei Pang, Lei Bai, Tianshu Yu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.23646v1 Announce Type: cross Abstract: Molecular embedding models can serve as foundational infrastructure for computational chemistry and drug discovery, where reusable vector representations support property prediction, virtual screening, and retrieval. Most molecular encoders are speci...
97. Scaling Reinforcement Learning for Diffusion Models via Velocity Matching ​
Author: Jaemoo Choi, Wei Guo, Yuchen Zhu, Arash Vahdat, Molei Tao, Julius Berner, Yongxin Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.23664v1 Announce Type: cross Abstract: Reward fine-tuning is becoming an important tool for adapting diffusion models to human preferences and task-specific objectives, but existing methods largely inherit policy-gradient machinery from large language models. Unlike autoregressive models,...
98. Automata from Agent Traces: Failure and Next-Step Prediction ​
Author: Seonglae Cho, Franklin Cardenoso Fernandez, Umar Mohammed, Zekun Wu, Kleyton Da Costa, Ilham Wicaksono, Adriano Koshiyama
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.23670v1 Announce Type: cross Abstract: LLM-based agents execute multi-step tasks, but their behavioral structure remains opaque: long unstructured traces resist the safety auditing and runtime monitoring that deployment requires. Existing approaches operate per-trace or success-only, so t...
99. A Hybrid Two-Stage Machine Learning Pipeline for Fault Detection and Classification in Power Transmission Systems ​
Author: Sahil Manikshete, Atharva Gujarathi, Thanh Long Vu, Akhtar Hussain, Van-Hai Bui
Published: 8/26/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2608.23726v1 Announce Type: cross Abstract: Rapid and accurate fault detection in high-voltage transmission networks is essential for grid reliability and equipment protection. Transmission fault datasets are frequently imbalanced, and certain fault types produce electrical signatures that fal...
100. S-matrix informed neural networks for amplitude analysis ​
Author: Wyatt A. Smith, Arkaitz Rodas, Marius D. Thomas, C'esar Fern'andez-Ram'irez, Giorgio Foti, Lin Qiu, Adam P. Szczepaniak, Alessandro Pilloni
Published: 8/26/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, nucl-th
arXiv:2608.23750v1 Announce Type: cross Abstract: Reconstructing scattering amplitudes from finite, noisy, and mutually inconsistent measurements is an ill-posed inverse problem common to many reactions relevant to particle physics. We introduce S-matrix informed neural networks (SINNs), and demonst...
101. (Mis)Understanding Benign Overfitting in Equity Return Prediction ​
Author: Hui Guo, Jiawei Huang, Runze Li, Yan Yu
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP, stat.ME
arXiv:2608.23761v1 Announce Type: cross Abstract: Highly overparameterized models often predict well despite interpolating training data in complex domains, challenging the classical bias--variance tradeoff. We investigate whether this ``benign overfitting'' phenomenon extends to equity return predi...
102. Accelerating the Adoption of Residential Solar Power Systems: Policy Analysis using a Dynamic Structural Model ​
Author: Sebasti'an Souyris, Jason A. Duan, Anantaram Balakrishnan, Varun Rai
Published: 8/26/2026, 4:00:00 AM
Categories: econ.EM, cs.LG, stat.AP
arXiv:2608.23796v1 Announce Type: cross Abstract: Problem definition: Solar electricity generation is a strategic component of energy portfolios designed to meet growing demand and reduce carbon emissions. Governments and municipalities encourage household photovoltaic (PV) adoption through upfront ...
103. Restoring Without Forgetting: Continual Learning Across Image Degradations ​
Author: Alif Ashrafee, Bartosz Krawczyk
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.23799v1 Announce Type: cross Abstract: Recent progress in image restoration has converged on all-in-one architectures that jointly handle multiple degradations within a single network. These methods are effective on static benchmarks but target a closed-world setting that assumes simultan...
104. LUCAID: Agentic Multimodal AI for Lung Cancer Precision Pathology ​
Author: Marie-Lisa Eich, Kai Standvoss, Timo Milbich, Alexander M"ollers, Miriam H"agele, Philipp Anders, Lars Tharun, Hanna Kontradiuk, Sebastian Kons, Nader Aldoj, Recepcan Adig"uzel, Adam Narai, Lukas H"onig, Jonathan Striebel, Binru Yang, Mihnea P. Dragomir, Marvin Sextro, Philipp Keyl, Philipp Jurmeister, Rosemarie Krupar, Evelyn Ramberger, James Wells, Julika Ribbat-Idel, Andreas Kunft, Hussam Shuaib, Christian Groh'e, Reinhard B"uttner, David Horst, Klaus-Robert M"uller, Lukas Ruff, Maximilian Alber, Frederick Klauschen, Simon Schallenberg
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.23803v1 Announce Type: cross Abstract: Lung cancer tissue diagnostics is complex, as therapy decisions in precision oncology rely on the integration of histomorphological, immunohistochemical, and molecular features. Yet pathological assessment remains largely visual and semi-quantitative...
105. A Formal Methodological Framework for Auditing Robustness and Fidelity in Explainable AI: From Application to Trust Certification ​
Author: Rosa Elysabeth Ralinirina, Jean Christian Ralaivao, Niaiko Micha"el Ralaivao, Alain Josu'e Ratovondrahona, Thomas Mahatody
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG
arXiv:2608.23817v1 Announce Type: cross Abstract: SHAP and LIME are now standard tools for interpreting black-box predictions, yet their outputs can vary substantially when the input is perturbed by small amounts of noise--a problem we observed firsthand in our previous work on food security in Mada...
106. Mitigating Exploration Bias in RL for Multi-Instruction Following ​
Author: Mian Zhang, Yueqin Yin, Kaiyu He, Peilin Wu, Xinlu Zhang, Mingyuan Zhou, Zhiyu Zoey Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.23830v1 Announce Type: cross Abstract: RL has emerged as a powerful paradigm for enhancing the instruction following capabilities of LLMs. While existing training recipes achieve substantial gains, we find that they suffer from exploration bias towards easy instructions when the training ...
107. Learning to Act While Waiting: RL Finetuning of Generalist Robot Policies Under Inference Latency ​
Author: Brian Zhu (Siemens), Momen Khalil (Siemens), E Harrison (UC Berkeley), Emanuele Poggi (Siemens), Philipp Schmitt (Siemens), Bernd Kast (Siemens), Philine Meister (Siemens), Pranav Atreya (UC Berkeley), Qiyang Li (UC Berkeley), Finn Ferchau (Siemens), Cesar Colmenero (Siemens), Yash Shahapurkar (Siemens), Gokul Narayanan (Siemens), Melih Erdogan (Siemens), Kai Wurm (Siemens), Georg von Wichert (Siemens), Oier Mees (Microsoft, ETH Zurich, UC Berkeley), Eugen Solowjow (Siemens), Andrew Wagenmaker (UC Berkeley), Sergey Levine (UC Berkeley)
Published: 8/26/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.23831v1 Announce Type: cross Abstract: While reinforcement learning (RL) allows generalist robot policies to continually improve during deployment, the large model size of modern generalist policies, such as VLAs, poses a fundamental obstacle to effective RL improvement. In particular, th...
108. Predicting Radiologist Expertise from 3D Gaze Patterns During CT Interpretation ​
Author: Leila Khaertdinova, Anna Anikina, Claudia Mello-Thoms, Bulat Ibragimov
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.23836v1 Announce Type: cross Abstract: Accurate interpretation of volumetric CT requires efficient navigation of 3D image volumes and attention to diagnostically relevant regions. While eye-tracking has been widely studied in 2D medical imaging, its use for expertise assessment in CT sett...
109. Infant Care Video Dataset for Classification of Interventions Using Transformers ​
Author: Igor Bogdanov, James Green
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.23838v1 Announce Type: cross Abstract: Healthcare documentation in the neonatal intensive care unit (NICU) presents significant challenges, with nurses spending approximately 25% of their time on record-keeping, while up to 60% of interventions remain undocumented. Motivated by the need...
110. Pipeline-Native Transformers: Co-Designing Model Architecture and CPU Inference for Bandwidth-Efficient Autoregressive Decode ​
Author: Tom Poperszky
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, cs.PF
arXiv:2608.23841v1 Announce Type: cross Abstract: Single-token autoregressive decode on CPUs is bound by memory bandwidth, not arithmetic: a modern CPU sustains roughly 1 TFLOP/s of compute but only about 50 GB/s from main memory, and each generated token must stream every active weight once. This r...
111. Exploit More, Explore Smarter for Budget-Constrained Agentic Search ​
Author: Haoyang Fang, Bernie Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.23848v1 Announce Type: cross Abstract: Budget-constrained agentic search arises when an LLM agent must refine candidates under a small evaluation budget, because validation is expensive, generation requires multiple model calls, or both. In this regime, standard MCTS allocates budget poor...
112. Semantic Overlays: Mitigating Prompt Injection with Annotations Beyond Tokens and Steering Vectors ​
Author: Joshua Penman
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.CR, cs.LG
arXiv:2608.23873v1 Announce Type: cross Abstract: Everything a language model sees is tokens. The serving stack knows what each span is -- user input, tool output, instructions -- but the model must keep track of that itself, and it can lose track or be confused: text can be written to read like any...
113. Differential Learning for Robust Prediction of Thermal Stability with Application to Energetic Materials ​
Author: Megan C. Davis, R. Seaton Ullberg, Jeremy N. Schroeder, Andrew H. Salij, Marc J. Cawkwell, Christopher J. Snyder, Ivana Matanovic, Wilton J. M. Kort-Kamp
Published: 8/26/2026, 4:00:00 AM
Categories: physics.chem-ph, cond-mat.mtrl-sci, cs.LG
arXiv:2608.23874v1 Announce Type: cross Abstract: Predicting thermal stability during handling and storage is essential for the design of safe and reliable energetic materials. However, experimental measurements vary significantly across laboratories due to differences in protocols and analysis meth...
114. Spatiotemporal Distillation via Recurrent Bottlenecks for Aortic Tracking ​
Author: Dexter Wen Jie Teo, Nairouz Shehata, Herve Lombaert
Published: 8/26/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2608.23879v1 Announce Type: cross Abstract: Cardiac cine-MRI serves as a direct visual indicator of cardiovascular hemodynamics by capturing the continuous wall motion of the aorta. Quantifying these dynamic structural changes across the cardiac cycle is essential for measuring aortic distensi...
115. Dimensionless Controls of Plasticity Under Alternating Tasks: From Evolutionary Biology to Continual Learning ​
Author: Owen Skriloff
Published: 8/26/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2608.23889v1 Announce Type: cross Abstract: Plasticity under changing environments is central to both evolutionary biology and continual learning. Motivated by recent work on genotype--phenotype maps, we study a minimal deep-learning analogue where a network is trained alternately on two Boole...
116. Provenance Guided Incremental Learning Under Evolving Concept Definitions ​
Author: Ismail Lamaakal
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.23893v1 Announce Type: cross Abstract: Learning systems deployed over long periods must adapt not only to statistical changes in incoming data, but also to revisions of the definitions that generate their prediction targets. Conventional concept-drift methods typically infer such changes ...
117. PROOF-Gen: From Optimized Data to Better Distillation ​
Author: Anh Ta, Junjie Zhu, Shahin Shayandeh
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.23911v1 Announce Type: cross Abstract: Supervised fine-tuning on teacher-generated trajectories is the standard first stage for distilling tool-calling capabilities into deployable models. Post-training pipelines that drive shipped tool-calling agents re-run this stage on a daily or weekl...
118. Generalization, memorization, and overfitting for diffusion models trained in the lazy high-dimensional regime ​
Author: Hugo Latourelle-Vigeant, Sinho Chewi, Aram-Alexandre Pooladian, John Sous, Theodor Misiakiewicz
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2608.23938v1 Announce Type: cross Abstract: Modern score-based generative models have achieved remarkable empirical success in high-dimensional tasks such as image, audio, and video synthesis. These models reduce distribution learning to a sequence of regression problems that, if solved exactl...
119. NeuronGuard: Robust LLM Safety Alignment via Ablation-Aware Safety Signal Redistribution ​
Author: Anjun Gao, Yueyang Quan, Yufei Xia, Zhuqing Liu, Minghong Fang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.IR, cs.LG
arXiv:2608.23959v1 Announce Type: cross Abstract: Safety alignment in large language models (LLMs) remains brittle against a growing spectrum of attacks. Jailbreak attacks bypass safety mechanisms through crafted prompts, while neuron-level attacks directly prune safety-critical neurons post-deploym...
120. RAGSentinel: Certifiable Geometric Consensus for Robust Retrieval-Augmented Generation ​
Author: Yueyang Quan, Anjun Gao, Yufei Xia, Minghong Fang, Zhuqing Liu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.IR, cs.LG
arXiv:2608.23965v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) improves the factuality of large language models by grounding responses in external documents, but it also exposes a critical security vulnerability: adversarial documents injected into the knowledge database can ...
121. Giraffe: A Mapping Architecture from Hidden Text Representations to Visual Embeddings for Efficient Graphic Design ​
Author: Nejla Ghaboosi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.23970v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have made significant progress in understanding and interpreting mul- timedia content. However, their ability to generate me- dia remains limited. Recent approaches have attempted to bridge this gap by transla...
122. Memory Is Not Always Needed: Characterizing Conditional Memory in Scientific Reasoning ​
Author: Zhen Bi, Xueshu Chen, Yan Wang, Zhizhi Peng, Haosen Hong, Zhen Wang, Zhixuan Chu, Bingyu Zhu, Jungang Lou
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.23982v1 Announce Type: cross Abstract: Scientific reasoning requires language models to retrieve specialized knowledge and incorporate it reliably into multi-step computation. Conditional memory provides an explicit lookup pathway that complements dense neural representations, but its use...
123. Hierarchical Skill Retrieval for Data-Efficient Adaptation of Vision-Language-Action Models ​
Author: Haoran Hao, Shahram Najam Syed, Jeff Schneider, Jeffrey Ichnowski
Published: 8/26/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2608.24042v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models pretrained on large-scale robot datasets provide a strong foundation for robot manipulation, their performance can degrade when adapted to new tasks with limited task-specific demonstrations. Retrieval offers...
124. Compression Trinity: Exploring Sparsity, Quantization, and Low-Rank Approximations for LLM Compression ​
Author: Mohammad Mozaffari
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.DC, cs.LG, cs.PF
arXiv:2608.24070v1 Announce Type: cross Abstract: Prohibitive computational and environmental costs impede the scalable deployment of Large Language Models (LLMs). Traditional compression techniques (sparsity, quantization, low-rank approximations) are typically applied in isolation, and each hits a...
125. RetrievalFormer: A Dual-Encoder Transformer for Efficient Approximate Nearest Neighbor Retrieval and Cold-Item Recommendation ​
Author: Theodore Rogers, Joe Standerfer, Dmitrii Timoshenko, Haoxue Li, Zuhaib Akhtar, Soyoung Yang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.24079v1 Announce Type: cross Abstract: A shared search-and-recommendation index must score new items from features alone because search has no exploration slot. In a public log covering both surfaces over one catalog, $38.6%$ of held-out query-search impressions show an item never previo...
126. Joint-Embedding Prediction of Masked Point Tubes for Self-Supervised Learning on 4D Point Cloud Videos ​
Author: Jheng-Ling Lee, Shang-Tse Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.24093v1 Announce Type: cross Abstract: Self-supervised representation learning for 4D point cloud videos is challenging because annotations are costly and reconstruction-based pretraining can overemphasize low-level geometric details. We propose a JEPA-style framework that learns from unl...
127. qshap: Fast Shapley Decomposition of $R^2$ for Gradient-Boosted Trees ​
Author: Zhongli Jiang, Min Zhang, Dabao Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.24104v1 Announce Type: cross Abstract: Numerous methods have been developed to quantify feature attributions in individual predictions for tree ensembles. However, many applications require global measures of feature contributions to overall model performance. Although local attribution s...
128. Anatomy of a Scam Call: What 10,000 real scam and spam calls reveal about how phone scammers operate ​
Author: Ethan Traister, Ankit Raj, Jiaqi Gan, Xingyu Shen, Tyler Wu, Yuchen Zhou, Tommy Duong, Kidus Zewde, Siying Chen, Simiao Ren
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CR, cs.CL, cs.CY, cs.LG
arXiv:2608.24127v1 Announce Type: cross Abstract: Telephone fraud is pervasive and costly, but its inner workings are rarely observed at scale. We analyze a complete corpus of 10,211 inbound scam and spam calls -- 913 hours of audio and 330,956 transcribed turns from 5,780 distinct numbers -- collec...
129. Preference Optimization for Non-Verbal Vocalization Synthesis ​
Author: Haoyang Li, Chenglin Xu, Junchuan Zhao, Yuang Cao, Liumeng Xue, Yiwen Guo, Eng Siong Chng
Published: 8/26/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.LG
arXiv:2608.24163v1 Announce Type: cross Abstract: Non-verbal vocalizations (NVs), such as laughter, coughs, and sighs, are essential for expressive TTS, but the effectiveness of preference optimization for NV generation remains poorly understood. We systematically study preference optimization for N...
130. Decoupling candidate dual AGN from chance superpositions in the GOTHIC survey via a deep-learning framework ​
Author: Bhavesh Mukheja, Snehanshu Saha, Anwesh Bhattacharya, Mousumi Das, Fran\c{c}oise Combes, Sudhanshu Barway
Published: 8/26/2026, 4:00:00 AM
Categories: astro-ph.GA, cs.CV, cs.LG, math-ph, math.MP, physics.data-an
arXiv:2608.24164v1 Announce Type: cross Abstract: Dual active galactic nuclei (DAGN) mark a critical phase in the evolution of merging galaxies and the pairing of supermassive black holes, yet they remain difficult to identify in large imaging surveys because of projection effects and limited spatia...
131. Paritok-4B: Intent-Conditioned Context Compression for Coding Agents ​
Author: Jiayu Shi, Luzhuo Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.SE
arXiv:2608.24188v1 Announce Type: cross Abstract: Coding agents re-send large file reads and tool outputs to a frontier LLM every turn, and this context dominates their token bill. General-purpose prompt compressors are trained on prose and suit code poorly: they paraphrase identifiers and drop the ...
132. A Heterogeneous Mixture of Experts Framework for Interpretable Machine Learning ​
Author: Soham Chatterjee, Rwitobroto Dey, Smarajit Bose
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.24195v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models provide a flexible framework for partitioning complex prediction problems into simpler local learning tasks through an input-dependent gating mechanism. Existing interpretable MoE approaches, such as Mixture of Decisio...
133. A Theory of Finite-Noise Optima and Generalization in Quantum Machine Learning ​
Author: Ziyu Zhang, Zikang Jia, Xiaosong Li, Yulong Dong
Published: 8/26/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.24229v1 Announce Type: cross Abstract: Quantum noise is expected to degrade quantum machine learning by driving circuits away from their noiseless implementations. Yet recent studies show moderate noise can reduce testing error, a behavior unexplained by weak-noise perturbative error accu...
134. Validation of HRV Studio: A Transparent and Quality-Control-Aware Platform for Heart Rate Variability Analysis ​
Author: Cyrus Mexon Evrard Djindot, Faliang Liu, Sylvain Laborde, Yinjia Zhang, Jessie Chen, Ming Li, Congrong Wang, Weixiong Rao, Qinpei Zhao
Published: 8/26/2026, 4:00:00 AM
Categories: physics.med-ph, cs.LG, cs.SE, q-bio.QM
arXiv:2608.24241v1 Announce Type: cross Abstract: Reproducibility of heart rate variability (HRV) analysis is limited by differences in preprocessing and computational conventions across software platforms. We developed HRV Studio, an open-source PyQt6-based desktop application integrating transpare...
135. Can a Dynamic Internal Field Govern a Transformer's Cognition? Certifiability, not Superiority, in Homeostatic Compute Control ​
Author: Francisco M. Arrabal-Campos, Ignacio Fernandez, Francisco G. Montoya, Alfredo Alcayde
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SY, eess.SY
arXiv:2608.24319v1 Announce Type: cross Abstract: An intelligent system does not merely reason: it governs its own reasoning - how much to compute, when to stop, which module to activate. Can that role be played by a dynamic internal field - a low-dimensional homeostatic state with explicit physics ...
136. Mind the Student: Behavioral and Contextual Cues for Automated Engagement Prediction in Online Learning ​
Author: Alperen Kantarci, Visvanathan Ramesh, Gemma Roig
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.HC, cs.LG
arXiv:2608.24340v1 Announce Type: cross Abstract: The prediction of student engagement from the online tutoring videos is difficult because engagement is a multidimensional construct comprising distinct behavioral, emotional, and cognitive states. A reliable prediction requires bringing together dif...
137. Sequential operator learning under dependent data ​
Author: Rafael Oliveira
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.24426v1 Announce Type: cross Abstract: Learning operators from sequentially collected data arises in adaptive experimental design, Bayesian optimization, and dynamical-system modelling, where observations may be dependent, and future inputs or sensing operators may depend on preceding dat...
138. Predictability of El Ni\~no from Delayed Observations ​
Author: Francisco J. Beron-Vera
Published: 8/26/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG, math.DS, nlin.CD
arXiv:2608.24428v1 Announce Type: cross Abstract: Using monthly Ni~no-3.4 anomalies through July 2026, we investigate how much predictive information is contained in delayed observations of the index. Ridge regression identifies informative delays, while multilayer perceptron and sparse identificat...
139. Low-Rank Ternary Adaptation for Fine-Tuning Transformers ​
Author: Alexandru-Dragos Manolache, Yunqiang Li, Jan van Gemert
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.24469v1 Announce Type: cross Abstract: Ternary transformers offer extreme memory and compute efficiency, but existing low-bit LoRA-based methods cannot directly fine-tune ternary weights. Current approaches either require dequantization, restoring low-bit base weights to higher precision ...
140. NeuralParker: A Reinforcement Learning Planner for Irregular Parking Environments ​
Author: Zihan Wang, Bai Huang, Yang Guan, Xiao Li, Haoyu Xu, Naizheng Wang, Shengbo Eben Li
Published: 8/26/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.24485v1 Announce Type: cross Abstract: Automated parking commonly assumes marked slots and short approach maneuvers. Delivery and service vehicles, however, may need to reach an operator-specified pose in an irregular bounded environment from a distant start. Existing learning-based parki...
141. SatDL: Jointly Optimizing Data Redistribution and Training for Satellite-Based Distributed Learning ​
Author: Hao Wu, Kin Whye Chew, Yizhan Han, Han Li, Jingxian Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2608.24516v1 Announce Type: cross Abstract: Satellite-based distributed learning promises to train machine-learning models directly in orbit using massive, globally dispersed sensor data, thereby avoiding large-scale data downloads to ground servers. However, training convergence is significan...
142. Provable Quantum--Classical Separation for Continuous Gibbs Sampling ​
Author: Enrico Olivucci, Mariia Sobchuk, Sehmimul Hoque, Jeffrey Hnybida, Kyungho W. Kim, Ala Shayeghi, Pooya Ronagh
Published: 8/26/2026, 4:00:00 AM
Categories: quant-ph, cs.DS, cs.ET, cs.LG
arXiv:2608.24527v1 Announce Type: cross Abstract: We prove the first quantum--classical separation for a sampling problem over a continuous domain. For a class of Gibbs states $p\propto e^{-\beta E}$ on the torus $\mathbb{T}^d$ with smooth ($s$-Gevrey) potential and barrier amplitude $\alpha=e^{\bet...
143. MoRF-AST: Calibrated Probabilistic Virtual Sensing for Structural Monitoring under Changing Operating Conditions ​
Author: Wingho Feng, Quanwang Li, Ming Zhong, Jingyu Yang, Chen Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CE, cs.LG
arXiv:2608.24531v1 Announce Type: cross Abstract: Probabilistic full-field reconstruction provides uncertainty-aware response evidence for structural reliability assessment, yet inference from sparse and noisy measurements remains underdetermined. Most existing methods overlook shifts between offlin...
144. $\texttt{findr}$: Transparent and Fair Credit Risk Decisions through Semi-Structured Regressions ​
Author: Victor Medina-Olivares, Stefan Lessmann, Jonathan Crook
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, q-fin.RM
arXiv:2608.24582v1 Announce Type: cross Abstract: Credit risk models increasingly need to combine predictive accuracy with transparent explanations and auditable fairness constraints. Logistic regression remains attractive because its coefficients are easy to interpret, but it can miss nonlinear str...
145. When Similarity Is Interaction-Driven: Quantum Kernels for Regime-Sensitive Learning ​
Author: Hanqiu Peng, Jianlong Lu, Ying Chen
Published: 8/26/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.24631v1 Announce Type: cross Abstract: Similarity in many decision systems is governed not by distance alone but by interactions among variables. In fraud and anomaly detection, small local perturbations can cross interaction-sensitive decision boundaries while leaving ambient distance al...
146. Maia 200: A Software Defined Dataflow System for Large-scale AI Acceleration ​
Author: Sherry Xu, Marco Heddes, Jackson Peng, Tom Savell, Monica Tang, Prashant Ranjan, Jesse Benson, Ofer Dekel, Saurabh Dighe, Anupama Kurpad, Artour Levin, Matthew Mattina, George Petre, Cheng Tang, Yuan Yu, Li Zhang, Torsten Hoefler
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AR, cs.AI, cs.DC, cs.ET, cs.LG
arXiv:2608.24664v1 Announce Type: cross Abstract: We introduce Maia 200, an advanced AI accelerator delivering high performance-10 145 Tflop/s FP4 and 5072 Tflop/s FP8 within a 750W TDP and 7 TB/s HBM bandwidth. Maia exemplifies a new class of Software Defined Locally Accessed Dataflow Architectures...
147. Confident at the moment of action: belief miscalibration in LLM play under hidden information ​
Author: Bhushan Kashinath Joshi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.24691v1 Announce Type: cross Abstract: Agentic systems increasingly gate actions on a model's own stated confidence, which assumes confidence tracks correctness at the moment of acting. We test this in a hidden-information chess variant where royal status can be secretly, repeatedly reloc...
148. Lifted Model Construction under Approximate Commutativity ​
Author: Malte Luttermann, Jan Speller, Tanya Braun, Marcel Gehrke, Ralf M"oller
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.24713v1 Announce Type: cross Abstract: Lifted inference algorithms enable scalable probabilistic inference even for large object domains by leveraging the indistinguishability of objects in a probability distribution. An essential prerequisite for constructing a lifted representation is t...
149. Method, Mind, and Morality: How People Make Sense of Artificial Intelligence ​
Author: Jacy Reese Anthis, Erik Brynjolfsson, James Evans
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CL, cs.LG, stat.ML
arXiv:2608.24748v1 Announce Type: cross Abstract: How can humans make sense of the rapid takeoff of artificial intelligence (AI)? We studied the sensemaking dynamics of AI through an open-ended, mixed-methods study with computational text analysis of millions of AI-related newspaper articles and soc...
150. Weakly Supervised Seafloor Segmentation for Seagrass Habitat Mapping in Side-Scan Sonar Imagery ​
Author: Hayat Rajani, Nuno Gracias, Rafael Garcia
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.24756v1 Announce Type: cross Abstract: Seagrass meadows are crucial blue-carbon habitats, and mapping their extent is a prerequisite for coastal management and carbon inventory. Optical satellite sensors cover large areas but cannot reach deep or turbid water, whereas side-scan sonar (SSS...
151. MoTE: Mixture of Task Experts for Multi-Task Video Understanding ​
Author: Muhammad Asad Ali, Umar Khan, Nadia Robertini, Didier Stricker
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.24763v1 Announce Type: cross Abstract: Procedural video-language models must solve heterogeneous tasks from the same visual evidence, including action recognition, forecasting, and procedure prediction. Dense transformer decoders share the same feed-forward networks across tasks, which ca...
152. Score-Based Ideal Observer Approximation via Denoising Score Matching for Signal-Known-Exactly Detection Tasks ​
Author: Weimin Zhou
Published: 8/26/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG, stat.CO
arXiv:2608.24768v1 Announce Type: cross Abstract: The Bayesian Ideal Observer (IO) establishes the theoretical upper bound on task performance for binary detection tasks. However, analytical computation of the IO test statistic is generally intractable. Numerical approaches based on Markov-chain Mon...
153. LAION-BVD: A 10-Million-Hour Open Video Dataset for Multimodal Pre-training ​
Author: Andreas Hochlehnert, Marianna Nezhurina, Mehdi Cherti, Andrej Radonjic, Thadd"aus Wiedemer, Christoph Schuhmann, Romain Beaumont, Wieland Brendel, Bernhard Sch"olkopf, A. Sophia Koepke, Jenia Jitsev, Matthias Bethge
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.24845v1 Announce Type: cross Abstract: We present LAION-BVD, a large-scale open video dataset for multimodal learning, which contains 1.3B platform-specific video URLs collected from CommonCrawl. From these, we download 80M videos with a total duration of 10 million hours. The dataset is ...
154. Parameterized Complexity of $L_p$-Lipschitz Constants for Input Convex Neural Networks and $L_p$-Norm Maximization over Zonotopes ​
Author: Aritra Das, Vincent Froese, Moritz Grillo, Debayan Gupta, Christoph Hertrich, Tharrshann Jayan Logarajah, Georg Loho, Mihir More, Moritz Stargalla
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CC, cs.DM, cs.LG, cs.NE
arXiv:2608.24865v1 Announce Type: cross Abstract: Lipschitz constants are a standard way to quantify the sensitivity of neural networks to small input perturbations, but computing them is difficult even for shallow ReLU networks. We study this problem for two-layer input-convex neural networks (ICNN...
155. What FID Hides: Detecting, Ranking, and Diagnosing Deviations in Generative Evaluation ​
Author: Hao Chen
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.24881v1 Announce Type: cross Abstract: Generative models are commonly ranked by Fr'echet Inception Distance (FID) and Kernel Inception Distance (KID), yet FID's first-two-moment summary can miss distributional differences, and a reported scalar gap alone is not a calibrated test against ...
156. Opponent Aware Reinforcement Learning ​
Author: Victor Gallego, Roi Naveiro, David Rios Insua, David Gomez-Ullate Oteiza
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:1908.08773v3 Announce Type: replace Abstract: In certain reinforcement learning (RL) scenarios there are adversaries trying to interfere with the underlying reward process for their own benefit. We introduce Threatened Markov Decision Processes (TMDPs) as a framework to support an agent agains...
157. AdAdaGrad: Adaptive Batch Size Schemes for Adaptive Gradient Methods ​
Author: Tim Tsz-Kit Lau, Han Liu, Mladen Kolar
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML
arXiv:2402.11215v4 Announce Type: replace Abstract: The choice of batch size in minibatch stochastic gradient optimization is critical for both optimization and generalization performance in large-scale model training. Although large-batch training is arguably the dominant paradigm in large-scale de...
158. Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation ​
Author: Madison Cooley, Shandian Zhe, Robert M. Kirby, Varun Shankar
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2406.02336v3 Announce Type: replace Abstract: We present polynomial-augmented neural networks (PANNs), a novel machine learning architecture that combines deep neural networks (DNNs) with polynomial expansions. PANNs combine the strengths of DNNs (flexibility and efficiency in higher-dimension...
159. Quantum Maximum Entropy Inference and Hamiltonian Learning ​
Author: Minbo Gao, Zhengfeng Ji, Fuchao Wei
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, quant-ph
arXiv:2407.11473v2 Announce Type: replace Abstract: Maximum entropy inference and learning of graphical models are pivotal tasks in learning theory and optimization. This work extends algorithms for these problems, including generalized iterative scaling (GIS) and gradient descent (GD), to the quant...
160. Focal Calibration Loss: Controlling Posterior Distortion in Deep Neural Classifiers ​
Author: Wenhao Liang, Liangwei Zheng, Wei Zhang, Weitong Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, stat.ML
arXiv:2410.18321v3 Announce Type: replace Abstract: Confidence calibration matters wherever a classifier's probabilities, not just its labels, are consumed downstream. We study Focal Calibration Loss (FCL), which adds a squared probability-error (multiclass Brier) anchor to the focal objective, $\ma...
161. QABBA: Error-Guaranteed Symbolic Time-Series Compression via Integer-Quantized Aggregation ​
Author: Erin Carson, Xinye Chen, Fei He, Cheng Kang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, stat.ML
arXiv:2411.15209v3 Announce Type: replace Abstract: The expansion of time-series data from sensors and monitoring systems has made compact representations increasingly important. Such representations should retain signal structure while cutting storage, transmission and computation costs. Adaptive B...
162. TLXML: Task-Level Explanation of Meta-Learning via Influence Functions ​
Author: Yoshihiro Mitsuka, Shadan Golestan, Zahin Sufiyan, Shotaro Miwa, Osmar R. Zaiane
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2501.14271v4 Announce Type: replace Abstract: Meta-learning enables models to rapidly adapt to new tasks by leveraging prior experience, but its adaptation mechanisms remain opaque, especially regarding how past training tasks influence future predictions. We introduce TLXML (Task-Level eXplan...
163. Stabilizing Temporal Difference Learning via Implicit Stochastic Recursion ​
Author: Hwanwoo Kim, Panos Toulis, Eric Laber
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, math.PR, stat.ML
arXiv:2505.01361v3 Announce Type: replace Abstract: Temporal difference (TD) learning is a foundational algorithm in reinforcement learning (RL). For nearly forty years, TD learning has served as a workhorse for applied RL as well as a building block for more complex and specialized algorithms. Howe...
164. Massive-STEPS: Massive Semantic Trajectories for Understanding POI Check-ins -- Dataset and Benchmarks ​
Author: Wilson Wongso, Hao Xue, Flora D. Salim
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2505.11239v4 Announce Type: replace Abstract: Understanding human mobility through Point-of-Interest (POI) trajectory modeling is increasingly important for applications such as urban planning, personalized services, and generative agent simulation. However, progress in this field is hindered ...
165. NeuronTune: Fine-Grained Neuron Modulation for Balanced Safety-Utility Alignment in LLMs ​
Author: Birong Pan, Jianhao Chen, Mayi Xu, Qiankun Pi, Yuanyuan Zhu, Ming Zhong, Tieyun Qian
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2508.09473v2 Announce Type: replace Abstract: Ensuring robust safety alignment while preserving utility is critical for the reliable deployment of Large Language Models (LLMs). However, current techniques fundamentally suffer from intertwined deficiencies: insufficient robustness against malic...
166. Round-trip Reinforcement Learning: Self-Consistent Training for Better Chemical LLMs ​
Author: Lecheng Kong, Xiyuan Wang, Yixin Chen, Muhan Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.01527v2 Announce Type: replace Abstract: Large Language Models (LLMs) are emerging as versatile foundation models for computational chemistry, handling bidirectional tasks like reaction prediction and retrosynthesis. However, these models often lack round-trip consistency. For instance, a...
167. Is the Hard-Label Cryptanalytic Model Extraction Really Polynomial? ​
Author: Akira Ito, Takayuki Miura, Yosuke Todo
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2510.06692v3 Announce Type: replace Abstract: Deep Neural Networks (DNNs) have attracted significant attention, and their internal models are now considered valuable intellectual assets. Extracting such a model via oracle access to a DNN is conceptually similar to extracting a secret key from ...
168. MolGA: Molecular Graph Adaptation with Pre-trained 2D Graph Encoder ​
Author: Xingtong Yu, Chang Zhou, Xinming Zhang, Yuan Fang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.07289v2 Announce Type: replace Abstract: Molecular graph representation learning is widely used in chemical and biomedical research. While pre-trained 2D graph encoders have demonstrated strong performance, they overlook the rich molecular domain knowledge associated with submolecular ins...
169. SketchGuard: Scaling Byzantine-Robust Decentralized Federated Learning via Sketch-Based Screening ​
Author: Murtaza Rangwala, Farag Azzedin, Richard O. Sinnott, Rajkumar Buyya
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2510.07922v5 Announce Type: replace Abstract: Byzantine-robust decentralized federated learning (DFL) protects peer-to-peer training from malicious clients. The dominant defenses rely on similarity-based filtering, in which each client exchanges full model vectors with every neighbor before an...
170. LTR-ICD: A Ranking-Aware Framework for Automatic ICD Coding ​
Author: Mohammad Mansoori, Amira Soliman, Farzaneh Etminani
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IR
arXiv:2510.13922v3 Announce Type: replace Abstract: Clinical notes contain unstructured text provided by clinicians during patient encounters. These notes are usually accompanied by a sequence of diagnostic codes following the International Classification of Diseases (ICD). Correctly assigning and o...
171. Monotone and Separable Set Functions: Characterizations and Neural Models ​
Author: Soutrik Sarangi, Yonatan Sverdlov, Nadav Dym, Abir De
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.23634v5 Announce Type: replace Abstract: Motivated by applications for set containment problems, we consider the following fundamental problem: can we design set-to-vector functions so that the natural partial order on sets is preserved, namely $S\subseteq T \text{ if and only if } F(S)\l...
172. Adaptive prediction theory combining offline and online learning ​
Author: Haizheng Li, Lei Guo
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2512.00342v2 Announce Type: replace Abstract: Real-world intelligence systems usually operate by combining offline learning and online adaptation with highly correlated and non-stationary system data or signals, which, however, has rarely been investigated theoretically in the literature. This...
173. Wait, Wait, Wait... Why Do Reasoning Models Loop? ​
Author: Charilaos Pipis, Shivam Garg, Vasilis Kontonis, Vaishnavi Shrivastava, Akshay Krishnamurthy, Dimitris Papailiopoulos
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.12895v2 Announce Type: replace Abstract: Reasoning models (e.g., DeepSeek-R1) generate long chains of thought to solve harder problems, but they often loop, repeating the same text at low temperatures or with greedy decoding. We study why this happens and what role temperature plays. With...
174. Towards Reproducibility in Predictive Process Mining: SPICE -- A Deep Learning Library ​
Author: Oliver Stritzel, Nick H"uhnerbein, Simon Rauch, Itzel Zarate, Lukas Fleischmann, Moike Buck, Attila Lischka, Christian Frey
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2512.16715v3 Announce Type: replace Abstract: In recent years, Predictive Process Mining (PPM) techniques based on artificial neural networks have evolved as a method for monitoring the future behavior of unfolding business processes and predicting Key Performance Indicators (KPIs). However, m...
175. QiMeng-ChipV-RTL: Exploiting Information Locality for IP-level Verilog Generation ​
Author: Hanqi Lyu, Di Huang, Yaoyu Zhu, Kangcheng Liu, Bohan Dou, Chongxiao Li, Pengwei Jin, Shuyao Cheng, Rui Zhang, Zidong Du, Qi Guo, Xing Hu, Yunji Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.00704v2 Announce Type: replace Abstract: The generation of Register-Transfer Level (RTL) code is a crucial yet labor-intensive step in digital hardware design, traditionally requiring engineers to manually translate complex specifications into thousands of lines of synthesizable Hardware ...
176. Anytime Pretraining: Horizon-Free Learning-Rate Schedules with Weight Averaging ​
Author: Alexandru Meterez, Pranav Ajit Nair, Depen Morwani, Cengiz Pehlevan, Sham Kakade
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC, stat.ML
arXiv:2602.03702v2 Announce Type: replace Abstract: Large language models are increasingly trained in continual or open-ended settings, where the total training horizon is not known in advance. Despite this, most existing pretraining recipes are not anytime: they rely on horizon-dependent learning r...
177. How to Achieve the Intended Aim of Deep Clustering Now, without Deep Learning ​
Author: Kai Ming Ting, Wei-Jie Xu, Hang Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.05749v2 Announce Type: replace Abstract: Deep clustering (DC) is often quoted to have a key advantage over $k$-means clustering. Yet, this advantage is often demonstrated using image datasets only, and it is unclear whether it addresses the fundamental limitations of $k$-means clustering....
178. ICA: Information-Aware Credit Assignment for Visually Grounded Long-Horizon Information-Seeking Agents ​
Author: Cong Pang, Xuyu Feng, Yujie Yi, Jiaqi Su, Zixuan Chen, Jiawei Hong, Tiankuo Yao, Nang Yuan, Jiapeng Luo, Lewei Lu, Xin Lou
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.10863v2 Announce Type: replace Abstract: Long-horizon reinforcement learning for information seeking agents remains difficult because terminal rewards reveal whether the final answer is correct, but not which acquired information enabled it. This difficulty is amplified by text-derived we...
179. You Can Learn Tokenization End-to-End with Reinforcement Learning ​
Author: Sam Dauncey, Roger Wattenhofer
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.13940v3 Announce Type: replace Abstract: Tokenization is a hardcoded compression step which remains in the training pipeline of Large Language Models (LLMs), despite a general trend towards architectures becoming increasingly end-to-end. Prior work has shown promising results at scale in ...
180. Topology enables learning-based hydrodynamic prediction of the global river system ​
Author: Hancheng Ren, Gang Zhao, Shuo Wang, Louise Slater, Dai Yamazaki, Shu Liu, Jingfang Fan, Xueying Li, Shibo Cui, Ziming Yu, Shengyu Kang, Depeng Zuo, Dingzhi Peng, Zongxue Xu, Bo Pang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, physics.geo-ph
arXiv:2602.22293v2 Announce Type: replace Abstract: Accurate river prediction is essential for water, food and energy security, yet remains challenging across entire river networks. Machine learning has transformed Earth-system modeling, but a system-level advance for river prediction lags for lack ...
181. Breaking the Tuning Barrier: Zero-Hyperparameters Yield Multi-Corner Analysis Via Learned Priors ​
Author: Wei W. Xing, Kaiqi Huang, Jiazhan Liu, Hong Qiu, Shan Shen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AR
arXiv:2603.13092v3 Announce Type: replace Abstract: Yield Multi-Corner Analysis validates circuits across 25+ Process-Voltage-Temperature corners, resulting in a combinatorial simulation cost of $O(K \times N)$ where $K$ denotes corners and $N$ exceeds $10^4$ samples per corner. Existing methods fac...
182. msData: A Millisecond-Resolution Network Dataset for Advancing Time Series Foundation Models ​
Author: Subina Khanal, Seshu Tirupathi, Merim Dzaferagic, Marco Ruffini, Torben Bach Pedersen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.16497v3 Announce Type: replace Abstract: Time series foundation models (TSFMs) require diverse, real-world datasets to adapt across varying domains and temporal frequencies. However, current large-scale datasets predominantly focus on low-frequency time series with sampling intervals, i.e...
183. A Bayesian Learning Approach for Drone Coverage Network: A Case Study on Cardiac Arrest in Scotland ​
Author: Tathagata Basu, Edoardo Patelli, Gianluca Filippi, Ben Parsonage, Christy Maddock, Massimiliano Vasile, Marco Fossati, Adam Loyd, Shaun Marshall, Paul Gowens
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, stat.AP
arXiv:2603.23134v2 Announce Type: replace Abstract: Drones are becoming popular as a complementary system for Emergency Medical Services (EMS). Although several pilot studies and flight trials have shown the feasibility of drone-assisted Automated External Defibrillator (AED) delivery, running a ful...
184. Model-Based Learning of Near-Optimal Finite-Window Policies in POMDPs ​
Author: Philip Jordan, Maryam Kamgarpour
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.01024v2 Announce Type: replace Abstract: We study model-based learning of finite-window policies in tabular partially observable Markov decision processes (POMDPs). A common approach to learning under partial observability is to approximate unbounded history dependencies using finite acti...
185. Learning from the Right Rollouts: Data Attribution for PPO-based LLM Post-Training ​
Author: Dong Shu, Denghui Zhang, Jessica Hullman
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.01597v2 Announce Type: replace Abstract: Traditional RL algorithms like Proximal Policy Optimization (PPO) typically train on the entire rollout buffer, operating under the assumption that all generated episodes provide a beneficial optimization signal. However, these episodes frequently ...
186. Test-Time Adaptation for EEG Foundation Models: A Systematic Study under Real-World Distribution Shifts ​
Author: Gabriel Jason Lee, Jathurshan Pradeepkumar, Jimeng Sun
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP
arXiv:2604.16926v3 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models have shown strong potential for learning generalizable representations from large-scale neural data, yet their clinical deployment is hindered by distribution shifts across clinical settings, devices, ...
187. Enhancing RL Generalizability in Robotics through SHAP Analysis of Algorithms and Hyperparameters ​
Author: Lingxiao Kong, Cong Yang, Oya Deniz Beyan, Zeyd Boukhers
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO
arXiv:2605.02867v3 Announce Type: replace Abstract: Despite significant advances in Reinforcement Learning (RL), model performance remains highly sensitive to algorithm and hyperparameter configurations, while generalization gaps across environments complicate real-world deployment. Although prior w...
188. $\alpha$-PFN: Fast Entropy Search via In-Context Learning ​
Author: Herilalaina Rakotoarison, Steven Adriaensen, Tom Viering, Carl Hvarfner, Samuel M"uller, Frank Hutter, Eytan Bakshy
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.07134v2 Announce Type: replace Abstract: Information-theoretic acquisition functions such as Entropy Search (ES) offer a principled exploration-exploitation framework for Bayesian optimization (BO). However, their practical implementation relies on complicated and slow approximations, i.e...
189. Lightweight Adaptive Feature Composition for Heterogeneous Downstream Adaptation of Wireless Foundation Models ​
Author: Yuxuan Shi, Tingting Yang, Li Sun, Liwen Jing, Kangning Ma, Yuwei Wang, Mengfan Zheng
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.10277v2 Announce Type: replace Abstract: Mobile systems increasingly rely on heterogeneous learning-enabled wireless functions, for which separate taskspecific models incur redundant training and model-management overhead. Wireless foundation models (WFMs) enable these functions to share ...
190. When Can One Neuron Fix Repetition Loops in LLMs? ​
Author: Aristotelis Lazaridis, Aman Sharma, Dylan Bates, Brian King, Vincent Lu, Jack FitzGerald
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.13705v2 Announce Type: replace Abstract: The Gemma 4 instruction-tuned models share a reproducible failure: on long factual enumeration prompts, such as TV episodes, the 88 IAU constellations, or the 151 original Pokemon, they collapse into repetition, either a tight verbatim loop or a li...
191. MortarBench: Evaluating Mortgage Loan Origination Agents ​
Author: Matthew Toles, Yunan Lu, Manav Munjal, Bojun Liu, Yuanhao Deng, Stephanie Selig, Derek Rindner, Cheng Li, Zhou Yu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.19416v3 Announce Type: replace Abstract: Loan origination is the process by which a lender creates a new loan, from application and underwriting through approval and funding. This process serves a critical role in evaluating the eligibility and level of risk posed by an applicant. Recentl...
192. TaLK: Text-attributed Graph Dataset Distillation via Coupling Language Model with Graph-Aware Kernel ​
Author: Yeongho Kim, Yeonje Choi, Kijung Shin
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.22975v2 Announce Type: replace Abstract: Text-attributed graphs (TAGs) are widely used in many real-world domains, and learning on TAGs requires jointly modeling text semantics and graph structure. A standard approach for modeling TAGs is to combine a language model (LM) and a graph neura...
193. A Unified Algebraic Framework for Classification Performance Evaluation ​
Author: Ronaldo C. Prati
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.04028v2 Announce Type: replace Abstract: We propose a unified algebraic framework for classification performance evaluation covering binary, multiclass, multilabel, ordinal, hierarchical, cost-sensitive, and soft-label settings. Actual and predicted labels are represented as binary indica...
194. Eluna: An Agentic LLM System for Automating Warehouse Operations with Reasoning and Task Execution ​
Author: Ning Liu, P Aditya Sreekar, Kalle Kujanp"a"a, Zhaoxuan Zhu, Kaiwen Liu, Chuanneng Sun, Jorge Marchena Menendez, Matthew Bales, Tianyu Yang, Shahnawaz Alam, Rose Yu, Baoyuan Liu, Kristina Klinkner, Shervin Malmasi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.08960v2 Announce Type: replace Abstract: Warehouse operations are governed by Standard Operating Procedures (SOPs) that encode complex, multi-system decision logic, which must be executed reliably under strict time constraints, yet LLM agents lack mechanisms to enforce procedural complian...
195. Application of machine learning to monster level prediction in tabletop RPG game design ​
Author: Jolanta 'Sliwa, Jakub Adamczyk
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.09196v3 Announce Type: replace Abstract: Designing balanced adversaries is a central but labor-intensive task in tabletop role-playing game (TTRPG) development. In systems such as Pathfinder, each monster is described by many numerical attributes that jointly determine its power, summariz...
196. Nonlinear Axiomatic Attribution for Cooperative Games ​
Author: Weida Li, Zhuanghua Liu, Yaoliang Yu, Bryan Kian Hsiang Low
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.09869v2 Announce Type: replace Abstract: The Shapley value is a widely used concept in attribution problems, as it uniquely satisfies the axioms of linearity, consistency, equal treatment, and efficiency. Often, the inclusion AUC metric is used to evaluate the quality of player rankings, ...
197. Discrete Diffusion Models: A Unified Framework from Tokenization to Generation ​
Author: Ye Yuan, Weien Li, Rui Song, Zeyu Li, Haochen Liu, Xiangyu Kong, Zixuan Dong, Linfeng Du, Zipeng Sun, Weixu Zhang, Jiaxin Huang, Changjiang Han, Yonghan Yang, Zichen Zhao, Xiuyuan Hu, Haolun Wu, Yankai Chen, Fengran Mo, Jikun Kang, Bowei He, Dawn Song, Philip S. Yu, Xue Liu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.13431v2 Announce Type: replace Abstract: Discrete denoising diffusion models (DDMs) have recently emerged as a compelling alternative to autoregressive (AR) modeling for discrete data, offering parallel generation and iterative global refinement capabilities. Unlike continuous diffusion, ...
198. Weak-to-Strong Learning in Decision Making ​
Author: Jingwei Ji, Renyuan Xu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18467v2 Announce Type: replace Abstract: Many operational decisions rely on predictive models that estimate uncertain outcomes conditional on observable contexts. Training such models, however, often faces a fundamental data asymmetry: labeled outcomes are scarce or costly to obtain, whil...
199. LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding ​
Author: Junsung Hwang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.24555v2 Announce Type: replace Abstract: Serving large language models at long context is bottlenecked by the key-value (KV) cache, which is read in full at every decode step. Attention keys are locally low-rank though globally high-rank: a fixed low-rank sketch shared across pages is pro...
200. When May a Model Replace the Experiment? Audits, Licenses, and the Price of Trust in Surrogate-Driven Design ​
Author: Shuangxiu (Max), Ma (Zachary), Wenhe (Zachary), Zhao
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.01378v2 Announce Type: replace Abstract: Design campaigns in chemistry, materials science, and machine learning share a bottleneck: determining how good a candidate truly is requires an expensive evaluation - an experiment, a first-principles simulation, or a full training run. Machine-le...
201. TabDPT-Turbo: Efficient In-Context Learning for Tabular Prediction ​
Author: Rasa Hosseinzadeh, Alex Labach, Zexin Xue, Shuyi Han, Valentin Thomas, Anthony L. Caterini
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.01400v2 Announce Type: replace Abstract: Tabular foundation models, driven by in-context learning, have rapidly grown in quality and popularity. However, recent approaches with either cell-based architectures or retrieval have sacrificed efficiency for raw performance, restricting their u...
202. Reproducible Evaluation of MoE Expert Caching: Replay Semantics, Workload Contamination, and Operating Regimes ​
Author: Yu Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.PF
arXiv:2608.07911v4 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache management an attractive lever: a policy that raised the hit rate would cut expert traffic per t...
203. An Efficient Minimax-Optimal Algorithm for Adversarial $m$-Set Bandits ​
Author: Francesco Bacchiocchi, Tommaso Cesari, Roberto Colomboni
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12231v2 Announce Type: replace Abstract: We study adversarial combinatorial bandits with $m$-set actions, where at each round the learner selects $m$ out of $d$ items and observes only the aggregate loss of the selected items. The resulting action set contains $K=\binom{d}{m}$ elements an...
204. The Impact of Temporal Context Length and Encoding Strategies on Self-Supervised ECG Representation Learning ​
Author: Ahmed Sameh, Ramzi Al-Sharawi, Yogatheesan Varatharajah
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, eess.SP
arXiv:2608.12695v2 Announce Type: replace Abstract: Self-supervised electrocardiogram (ECG) models are often trained on a few seconds of ECG signal and, increasingly, on discretized token sequences. It remains unclear whether these choices sacrifice information needed for rhythm inference and longit...
205. ER-KANs: Efficient and Robust Kolmogorov-Arnold Networks for Data-Scarce Scientific Machine Learning ​
Author: Harshil Lodhiya
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.14773v2 Announce Type: replace Abstract: The efficient-KAN literature---covering Chebyshev, wavelet, and radial-basis-function variants of the original Kolmogorov-Arnold Network---has been benchmarked almost entirely on clean data. We show that this choice conceals a large capability diff...
206. MAPLE: MoE Adaptive Plug-and-play Layer-wise Expert allocation ​
Author: Lie Li, Wen Li, Junxiao Shen, Guosheng Hu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.15299v2 Announce Type: replace Abstract: Sparsely-activated Mixture-of-Experts (MoE) Transformers universally fix the same number of routed experts across all layers, a convention that ignores the well-documented heterogeneity in layer-wise redundancy. We demonstrate that this uniformity ...
207. GEAR: Generative Expansion and Real Anchoring for Two-Stage Distillation of Tabular Foundation Models ​
Author: Qi Qin, Jiajie Zhu, Dali Chen, Yuzhao Zhang, Jia-Xing Han, Peng Zhang, Ying Yan, Yifan Sun, Yu Su
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML
arXiv:2608.18849v2 Announce Type: replace Abstract: Tabular foundation models (TFMs) achieve strong performance through in-context learning, but context-dependent inference imposes substantial latency and memory costs, hindering large-scale deployment. We propose GEAR (\emph{Generative Expansion and...
208. Credit Without Ground Truth: Auditing Step-Level Credit Assignment in LLM Agents Against Executed Replay ​
Author: Haiyue Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.19760v2 Announce Type: replace Abstract: Audited against policy-conditional ground truth from executed replay in a single-agent tool environment (ALFWorld), none of the step-level credit signals we audit -- LLM-judge scores, outcome-conditioned logprob ratios, or the policy's own confiden...
209. Multi-Source Complex Network Reconstruction via Wasserstein Distributionally Robust Optimization and Algorithm Unrolling ​
Author: Chuansen Peng, Yifan Xia, Jinshan Zhong, Xiaojing Shen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.19914v2 Announce Type: replace Abstract: Reconstructing complex network topologies from data is a fundamental challenge in cybernetics and graph signal processing, with applications in neuroscience, sensor, and social networks. In practice, target-domain samples are scarce while heterogen...
210. Metag: A dataset to build agentic meta-reviewing capabilities ​
Author: Anirudh Sundar, Min Chen, Divya Tadimeti, Gemma Zhang, Xinyi Alice Li, Nigel Boachie Kumankumah, Pavan Uttej Ravva, Sadid Hasan, Somya Chatterjee, Pruthvi Prakash Navada, Xiao Wang, Yue Kang, Sulaiman Vesal, Larry Heck
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.20488v2 Announce Type: replace Abstract: AI tools increasingly support tasks across the scientific research cycle, from experiment design and manuscript preparation to peer review. At the same time, the continuing growth in conference submissions has increased the burden on meta-reviewers...
211. Scaling Muon for Diffusion Transformers ​
Author: Chenghao Li, Xiao Han, Xinxin Huang, Wei Liu, Boyang Li, Bing Xiao, Heran Zhang, Juanma Perez Rua, Ke Xu, Kangning Liu, Linjun Kuang, Na Li, Tan Wang, Tian Xie, Wei Peng, Yang Pei, Yifan Xu, Yuanhao Zhai, Yuwei Lin, Zhe Wang, Zihao He, Daniel Li, Junbiao Tang, Ziyang Jiang, Dake Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.20818v2 Announce Type: replace Abstract: The matrix-aware optimizer Muon improves large model training by balancing updates across singular directions, yet its scaling behavior and end-to-end efficiency on large Diffusion Transformers (DiTs) remain unclear. We first establish Muon's scali...
212. Across-Design Uncertainty in Short Pricing Panels: Inference and Identification ​
Author: Pedro Cadahia Delgado
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, econ.EM
arXiv:2608.21334v2 Announce Type: replace Abstract: Short observational pricing panels often contain many observations but few distinct price movements. We evaluate the inferential consequences of this sparsity in a synthetic data-generating process by separating estimation error into uncertainty co...
213. Blockwise Stabilized Adaptive Cubic Regularization with Subsolvers via Recurrence ​
Author: Rodion Podorozhny
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA
arXiv:2608.22129v2 Announce Type: replace Abstract: Cubic regularized Newton methods have the optimal $\mathcal{O}(\epsilon^{-3/2})$ global rate, but a dense subproblem solve limits the feasible block size. Scalable Cubic Newton variants replace the true block curvature with a diagonal, low-rank, Kr...
214. SANE: State Anomaly Neutralization for Stable Extreme-Context Delta-Rule Models ​
Author: Qingwen Lin, Boyan Xu, Xiao Liu, Zhifeng Hao, Ruichu Cai
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.22354v2 Announce Type: replace Abstract: Delta-Rule recurrent models maintain a fixed-size state, enabling $O(1)$ inference memory but potentially becoming unstable under extreme-context extrapolation. By tracking RWKV-7 over sequences of up to 100M tokens, we empirically identify a disti...
215. Functional compatibility as a determinant of persistent neural learning ​
Author: Hossein Javidnia
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.22462v2 Announce Type: replace Abstract: Neural networks can acquire new capabilities while damaging existing ones, but what determines whether new learning persists remains unclear. We identify functional compatibility, the extent to which incoming learning can coexist with behaviour tha...
216. Change Detection in Probability Flow ODE: Online Testing in Diffusion Latent Spaces ​
Author: Artem Kraevskiy, Artem Prokhorov
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.22807v2 Announce Type: replace Abstract: A rapidly growing range of sequential data tasks, such as identifying trend reversals in financial markets, auto-segmenting video and audio recordings, detecting changes in movement direction from motion sensors cannot be fully addressed without de...
217. The Mask Is Not the Model: Auditing Prefix Invariance in Attention, State-Space, and Hybrid Sequence Models ​
Author: Taebong Kim, Youngsik Hong, Minsik Kim, Sunyoung Choi, Jaewon Jang, Minseo Kim
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.22876v2 Announce Type: replace Abstract: Hybrid sequence models must satisfy prefix invariance: representations at position t must not depend on future inputs, yet this is rarely verified. We formalize prefix invariance and give a lightweight audit, two forward passes, no training or grad...
218. How Much Regularization Survives Averaging? Update Masking in Federated Learning ​
Author: Wenhao Yan, Fu Kuroda, Yucheng Jin, Zhenke Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.23286v2 Announce Type: replace Abstract: Federated learning on non-IID data seeks flat minima to generalize across clients, and existing methods borrow sharpness-aware minimization from centralized training. There is a second way to reach flat minima, in which the regularization comes for...
219. Spectrum-Aware Bounds on Invertibility for Privacy-Enhancing Instance Encoding ​
Author: Seokjin Hwang, Yuting Li, Kiwan Maeng
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2608.23382v2 Announce Type: replace Abstract: Instance encoding is a popular empirical technique for privacy enhancement when sharing data to an untrusted server. It transforms sensitive data through an encoding process before sharing, with the hope that the encoding process retains utility bu...
220. Best Practice Critic Optimization ​
Author: Penghui Qi, Xiangxin Zhou, Wee Sun Lee
Published: 8/26/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.23566v2 Announce Type: replace Abstract: Group-based reinforcement learning methods such as GRPO for large language models avoid training a critic by sampling multiple responses for each prompt. A reliable critic could instead estimate token-level advantages from one response, but standar...
221. A Discriminative Latent-Variable Model for Bilingual Lexicon Induction ​
Author: Sebastian Ruder, Ryan Cotterell, Yova Kementchedjhieva, Anders S{\o}gaard
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, stat.ML
arXiv:1808.09334v4 Announce Type: replace-cross Abstract: We introduce a novel discriminative latent variable model for bilingual lexicon induction. Our model combines the bipartite matching dictionary prior of Haghighi et al. (2008) with a representation-based approach (Artetxe et al., 2017). To tr...
222. Deep Feature Pyramid Convolutional Networks with In-Place Activated Batch Normalization for Automated Skin Lesion Boundary Segmentation ​
Author: Glib Kechyn
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, stat.ML
arXiv:1812.00877v2 Announce Type: replace-cross Abstract: Segmentation of skin lesion boundaries in dermoscopic imaging is an important prerequisite step for computer-aided diagnosis of malignant melanoma, but remains challenging due to fuzzy margins, occluding artifacts such as hair and blood vesse...
223. Machine Learning Classification and Portfolio Construction: Does the Loss Function Matter? ​
Author: Yang Bai, Kuntara Pukthuanthong
Published: 8/26/2026, 4:00:00 AM
Categories: q-fin.GN, cs.LG, econ.GN, q-fin.CP, q-fin.EC, q-fin.PM
arXiv:2108.02283v5 Announce Type: replace-cross Abstract: Classification outperforms regression across matched machine learning models in portfolio construction. A stacking ensemble of gradient boosted tree, random forest, and neural network yields a value-weighted annualized Sharpe ratio of 1.83 fo...
224. Nonconvex-Nonconcave Min-Max Optimization with a Small Maximization Domain ​
Author: Dmitrii M. Ostrovskii, Babak Barazandeh, Meisam Razaviyayn
Published: 8/26/2026, 4:00:00 AM
Categories: math.OC, cs.GT, cs.LG
arXiv:2110.03950v3 Announce Type: replace-cross Abstract: We study the problem of finding approximate first-order stationary points in optimization problems of the form $\min_{x \in X} \max_{y \in Y} f(x,y)$, where the sets $X,Y$ are convex and $Y$ is compact. The objective function $f$ is smooth, b...
225. RACR-MIL: Rank-aware contextual reasoning for weakly supervised grading of squamous cell carcinoma using whole slide images ​
Author: Anirudh Choudhary, Mosbah Aouad, Krishnakant Saboo, Angelina Hwang, Jacob Kechter, Blake Bordeaux, Puneet Bhullar, David DiCaudo, Steven Nelson, Nneka Comfere, Emma Johnson, Olayemi Sokumbi, Jason Sluzevich, Leah Swanson, Dennis Murphree, Aaron Mangold, Ravishankar Iyer
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2308.15618v3 Announce Type: replace-cross Abstract: Squamous cell carcinoma (SCC) is one of the most common cancer subtype, with an increasing incidence and a significant impact on cancer-related mortality. SCC grading using whole slide images is inherently challenging due to the lack of a sta...
226. GNNBleed: Inference Attacks to Unveil Private Edges in Graphs with Realistic Access to GNN Models ​
Author: Zeyu Song, Ehsanul Kabir, Shagufta Mehnaz
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2311.16139v3 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) have become indispensable tools for learning from graph structured data, catering to various applications such as social network analysis and fraud detection for financial services. At the heart of these networks ...
227. HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA ​
Author: Xinyue Chen, Pengyu Gao, Jiangjiang Song, Xinjian Chen, Xiaoyang Tan
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2402.01767v4 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) significantly improves document-based question answering by integrating external documents during generation. However, retrieval accuracy can degrade when the knowledge base contains many semantically and ...
228. LEMMA-RCA: A Large Multi-modal Multi-domain Dataset for Root Cause Analysis ​
Author: Lecheng Zheng, Zhengzhang Chen, Dongjie Wang, Chengyuan Deng, Reon Matsuoka, Haifeng Chen
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2406.05375v4 Announce Type: replace-cross Abstract: Root cause analysis (RCA) is crucial for enhancing the reliability and performance of complex systems. However, progress in this field has been hindered by the lack of large-scale, open-source datasets tailored for RCA. To bridge this gap, we...
229. Intrinsic PAPR: Tackling Misattribution in 3D Intrinsic Decomposition via Proximity Attention Point Rendering ​
Author: Alireza Moazeni, Shichong Peng, Yanshu Zhang, Chirag Vashist, Ke Li
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.GR, cs.LG
arXiv:2407.00500v2 Announce Type: replace-cross Abstract: Recent point-based intrinsic decomposition and inverse rendering methods have advanced the modelling of the shading and albedo of 3D scenes. However, we identify a fundamental limitation: these methods suffer from a misattribution issue, wher...
230. Two-Sided Nearest Neighbors: An adaptive and minimax optimal procedure for matrix completion ​
Author: Tathagata Sadhukhan, Manit Paul, Raaz Dwivedi
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.ME, stat.TH
arXiv:2411.12965v3 Announce Type: replace-cross Abstract: Nearest neighbor (NN) algorithms have been extensively used for missing data problems in recommender systems and sequential decision-making systems. Prior theoretical analysis has established favorable guarantees for NN when the underlying da...
231. Highway Congestion Reduction through Reinforcement Learning Based Eulerian Headway Control ​
Author: Yaron Veksler, Sharon Hornstein, Han Wang, Maria Laura Delle Monache, Daniel Urieli
Published: 8/26/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.LG, cs.SY, eess.SY
arXiv:2412.02520v4 Announce Type: replace-cross Abstract: Connected automated vehicles (CAVs) equipped with adaptive cruise control (ACC) create new opportunities for highway congestion mitigation. Traditional practice relies on Eulerian variable speed limits (VSL) which regulate traffic through roa...
232. Contextual Online Uncertainty-Aware Preference Learning for Human Feedback ​
Author: Nan Lu, Ethan Lee, Ethan X. Fang, Junwei Lu
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2504.19342v4 Announce Type: replace-cross Abstract: Reinforcement Learning from Human Feedback (RLHF) has become a pivotal paradigm in artificial intelligence to align large models with human preferences. In this paper, we propose a novel statistical framework to simultaneously conduct the onl...
233. Improved generalization bounds for binary linear classification via isoperimetry ​
Author: Shogo Nakakita
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2505.16713v4 Announce Type: replace-cross Abstract: We examine the concentration of uniform generalization errors around their expectation in binary linear classification problems via an isoperimetric argument. In particular, we establish Poincar'{e} and log-Sobolev inequalities for the joint...
234. Asymptotically perfect seeded graph matching without edge correlation (and applications to inference) ​
Author: Tong Qi, Vera Andersson, Peter Viechnicki, Vince Lyzinski
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2506.02825v3 Announce Type: replace-cross Abstract: We present the OmniMatch algorithm for seeded multiple graph matching. In the setting of $d$-dimensional Random Dot Product Graphs (RDPG), we prove that under mild assumptions, OmniMatch with $s$ seeds asymptotically and efficiently perfectly...
235. Quasar: A Programming Language Specialized for LLM Code Actions ​
Author: Stephen Mell, Botong Zhang, David Mell, Shuo Li, Ramya Ramalingam, Nathan Yu, Stephan Zdancewic, Osbert Bastani
Published: 8/26/2026, 4:00:00 AM
Categories: cs.PL, cs.AI, cs.CR, cs.LG
arXiv:2506.12202v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often call external tools to solve tasks. One effective strategy is for LLMs to write code, enabling them to use complex control flow such as conditionals and loops. Such code actions are typically represented as ...
236. A Modular Multitask Reasoning Framework Integrating Spatio-temporal Models and LLMs ​
Author: Kethmi Hirushini Hettige, Jiahao Ji, Cheng Long, Shili Xiang, Gao Cong, Jingyuan Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2506.20073v3 Announce Type: replace-cross Abstract: Spatio-temporal data mining plays a pivotal role in informed decision making across diverse domains. However, existing models are often restricted to narrow tasks, lacking the capacity for multi-task inference and complex long-form reasoning ...
237. Iwin Transformer: Hierarchical Vision Transformer using Interleaved Windows ​
Author: Simin Huo, Ning Li
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2507.18405v3 Announce Type: replace-cross Abstract: Vision Transformers (ViTs) face two limitations: the rigid resolution dependency of positional embeddings, which complicates cross-resolution fine-tuning, and the quadratic complexity of attention. While Swin Transformer alleviates the latter...
238. Adaptive Multi-Mode Out-of-Distribution Detection for Trajectory Prediction in Autonomous Vehicles ​
Author: Tongfei Guo, Lili Su
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO
arXiv:2509.13577v3 Announce Type: replace-cross Abstract: Trustworthy trajectory prediction grounds autonomous vehicle (AV) safety, yet deployed models inevitably face out-of-distribution (OOD) scenes. Prior AV OOD detection targets perception, but planners act on predicted futures rather than raw s...
239. Reconquering Bell sampling on qudits: stabilizer learning and testing, quantum pseudorandomness bounds, and more ​
Author: Jonathan Allcock, Joao F. Doriguello, G'abor Ivanyos, Miklos Santha
Published: 8/26/2026, 4:00:00 AM
Categories: quant-ph, cs.CC, cs.DS, cs.LG
arXiv:2510.06848v3 Announce Type: replace-cross Abstract: Bell sampling is a simple yet powerful tool based on measuring two copies of a quantum state in the Bell basis, and has found applications in a plethora of problems related to stabiliser states and measures of magic. However, it was not known...
240. CytoNet: A Foundation Model for the Human Cerebral Cortex at Cellular Resolution ​
Author: Christian Schiffer, Zeynep Boztoprak, Jan-Oliver Kropp, Julia Th"onni{\ss}en, Katia Berr, Hannah Spitzer, Mathis Bode, Thomas Lippert, Katrin Amunts, Timo Dickscheid
Published: 8/26/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.LG
arXiv:2511.01870v3 Announce Type: replace-cross Abstract: Studying the cellular architecture of the human cerebral cortex is essential for understanding how the brain is organized from the micro to the macro level, and how it functions. However, investigating complex texture patterns in histological...
241. A Robust Task-Level Control Architecture for Learned Dynamical Systems ​
Author: Eshika Pathak, Ahmed Aboudonia, Sandeep Banik, Naira Hovakimyan
Published: 8/26/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SY, eess.SY
arXiv:2511.09790v2 Announce Type: replace-cross Abstract: Dynamical system (DS)-based learning from demonstration (LfD) is a powerful tool for generating motion plans in the operation ('task') space of robotic systems. However, realizing generated motion plans is often compromised by a "task-executi...
242. Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations ​
Author: Qianli Wang, Nils Feldhus, Pepa Atanasova, Fedor Splitt, Simon Ostermann, Sebastian M"oller, Vera Schmitt
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2601.00282v2 Announce Type: replace-cross Abstract: Quantization is widely used to accelerate inference and streamline the deployment of large language models (LLMs), yet its effects on self-explanations (SEs) remain unexplored. SEs, generated by LLMs to justify their own outputs, require reas...
243. E2HiL: Entropy-Guided Sample Selection for Efficient Real-World Human-in-the-Loop Reinforcement Learning ​
Author: Haoyuan Deng, Yudong Lin, Yuanjiang Xue, Haoyang Du, Qianzhun Wang, Boyang Zhou, Zhenyu Wu, Ziwei Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2601.19969v2 Announce Type: replace-cross Abstract: Human-in-the-loop guidance has emerged as an effective approach for accelerating online reinforcement learning (RL) in real-world manipulation. However, existing human-in-the-loop RL (HiL-RL) frameworks often suffer from low sample efficiency...
244. Comparing Explanations is Not Enough, Explain the Change: New Standards are Needed to Explain Behavioral Shifts in Large Language Models ​
Author: Martino Ciaperoni, Marzio Di Vece, Roberto Pellungrini, Luca Pappalardo, Fosca Giannotti, Francesco Giannini
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2602.02304v3 Announce Type: replace-cross Abstract: Large-scale foundation models exhibit behavioral shifts when subjected to interventions such as scaling, fine-tuning, reinforcement learning with human feedback, or in-context learning. Current explainability methods are structurally ill-suit...
245. MPIB: A Benchmark for Medical Prompt Injection Attacks and Clinical Safety in LLMs ​
Author: Junhyeok Lee, Han Jang, Kyu Sung Choi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2602.06268v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) systems are increasingly integrated into clinical workflows. However, prompt injection attacks can steer these systems toward clinically unsafe or misleading outputs. We in...
246. Do physics-informed neural networks (PINNs) need to be deep? Shallow PINNs using the Levenberg-Marquardt algorithm ​
Author: Muhammad Luthfi Shahab, Imam Mukhlash, Hadi Susanto
Published: 8/26/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, cs.NE, math.OC
arXiv:2602.08515v3 Announce Type: replace-cross Abstract: This work investigates shallow physics-informed neural networks (PINNs) for solving forward and inverse problems governed by nonlinear partial differential equations (PDEs). By formulating PINN training as a nonlinear least-squares problem, t...
247. ST-Lite: Training-Free KV Cache Compression with Spatio-Trajectory Guidance for Long-Horizon GUI Agents ​
Author: Bowen Zhou, Zhou Xu, Wanli Li, Jingyu Xiao, Pingan Gan, Haoqian Wang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2603.00188v2 Announce Type: replace-cross Abstract: Training-free KV cache compression is essential for deploying vision-language GUI agents under memory and latency constraints, yet existing methods are designed for generic language workloads and ignore the distinctive structure of GUI intera...
248. Bayes with No Shame: Admissibility Geometries of Predictive Inference ​
Author: Nicholas G. Polson, Daniel Zantedeschi
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2603.05335v3 Announce Type: replace-cross Abstract: Modern predictive systems combine predictors, sequential monitors, prediction sets, and online strategies, each with a different certificate of optimality. We study four criterion-relative geometries: Blackwell risk dominance, anytime-valid a...
249. Holographic Invariant Storage: Design-Time Safety Contracts via Vector Symbolic Architectures ​
Author: Arsenios Scrivens
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.CL, cs.IT, cs.LG, math.IT
arXiv:2603.13558v2 Announce Type: replace-cross Abstract: We introduce Holographic Invariant Storage (HIS), a protocol that assembles known properties of bipolar Vector Symbolic Architectures into a design-time safety contract for LLM context-drift mitigation. The contract provides three closed-form...
250. Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models ​
Author: Xiaojie Gu, Sherry T. Tong, Aosong Feng, Sophia Simeng Han, Jinghui Lu, Yingjian Chen, Yusuke Iwasawa, Yutaka Matsuo, Chanjun Park, Rex Ying, Irene Li
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2603.16654v3 Announce Type: replace-cross Abstract: Evaluating the reasoning abilities of large language models (LLMs) solely from final answers can obscure failures in intermediate steps, especially in multi-hop QA benchmarks without step-level annotations. To address this gap, we introduce O...
251. Lightweight GenAI for Network Traffic Generation: Fidelity, Augmentation, and Classification ​
Author: Giampaolo Bovenzi, Domenico Ciuonzo, Jonatan Krolikowski, Antonio Montieri, Alfredo Nascita, Antonio Pescap`e, Dario Rossi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.NI, cs.AI, cs.LG
arXiv:2603.25507v2 Announce Type: replace-cross Abstract: Network Traffic Classification (NTC) increasingly relies on data-driven models, yet its practical deployment is often constrained by limited labeled data, strict privacy requirements, and the cost of collecting representative traffic traces. ...
252. NAIMA: Semantics Aware RGB Guided Depth Super-Resolution ​
Author: Tayyab Nasir, Daochang Liu, Ajmal Mian
Published: 8/26/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, cs.MM
arXiv:2604.04407v2 Announce Type: replace-cross Abstract: Guided depth super-resolution (GDSR) is a multi-modal approach for depth map super-resolution that relies on a low-resolution depth map and a high-resolution RGB image to restore finer structural details. However, the misleading color and tex...
253. The Theorems of Dr. David Blackwell and Their Contributions to Artificial Intelligence ​
Author: Napoleon Paxton
Published: 8/26/2026, 4:00:00 AM
Categories: cs.GL, cs.LG, stat.ML
arXiv:2604.06621v2 Announce Type: replace-cross Abstract: Dr. David Blackwell was a mathematician and statistician of the first rank, whose contributions to statistical theory, game theory, and decision theory predated many of the algorithmic breakthroughs that define modern artificial intelligence....
254. Contextual Memory-Enhanced Source Coding for Low-SNR Communications ​
Author: Ziqiong Wang, Rongpeng Li, Zhifeng Zhao, Honggang Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT
arXiv:2605.04400v3 Announce Type: replace-cross Abstract: Separate Source-Channel Coding (SSCC) remains vulnerable in noisy text transmission due to the fragility of autoregressive source decoding, especially when Arithmetic Coding (AC) relies on Large Language Model (LLM)-based probability estimati...
255. Superintelligent Retrieval Agent: The Next Frontier of Agentic Retrieval ​
Author: Zeyu Yang, Xu Han, Qi Ma, Jason Chen, Anshumali Shrivastava
Published: 8/26/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG
arXiv:2605.06647v3 Announce Type: replace-cross Abstract: Retrieval-augmented agents are increasingly the interface to large knowledge bases, yet most treat retrieval as a black box: they issue exploratory queries, inspect snippets, and reformulate until evidence emerges. This resembles how a newcom...
256. EngiAI: Capability-Based Evaluation of Tool-Connected LLM Agents for Engineering Design ​
Author: Gioele Molinari, Florian Felten, Soheyl Massoudi, Mark Fuge
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA
arXiv:2605.19743v3 Announce Type: replace-cross Abstract: Engineering-agent systems are proliferating, but differences in tasks, tools, and success criteria make demonstrations difficult to compare and failures difficult to diagnose. We introduce a capability-based evaluation framework for tool-conn...
257. Encrypted Neural Networks without Overflows ​
Author: Philipp Kern, Lorenzo Rovida, Samuel Teuber, Edoardo Manino, Carsten Sinz, Alberto Leporati
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2605.23096v2 Announce Type: replace-cross Abstract: The popular Cheon-Kim-Kim-Song (CKKS) scheme enables efficient private inference in neural networks by evaluating them on encrypted data. Since CKKS only supports addition, multiplication, and array rotation operations, turning neural network...
258. From Local Geometry to Global Pseudo Labeling for Robust Positive Unlabeled Learning under Covariate Shift ​
Author: Firas Gabetni, Alexandre Rocchi, Nacim Belkhir, Ziyi Liu, Gianni Franchi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2605.31187v2 Announce Type: replace-cross Abstract: Detecting covariate shift is critical for building reliable vision systems. While most prior work focuses on improving robustness to shift, explicitly detecting covariate shift remains underexplored. Existing approaches typically rely on full...
259. A Circuit, Not The Circuit: Non-Unique Causal Localisation of the Mamba-2 State Sink ​
Author: Yuhang Jiang, Bowen Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2606.00930v2 Announce Type: replace-cross Abstract: Mechanistic interpretability routinely reads a probe and labels its top-activating units as the circuit executing the computation. We test the move in Mamba, on the state sink: the selective state-space analogue of the Transformer attention s...
260. An Algebraic View of the Expressivity of Recurrent Language Models ​
Author: Franz Nowak, Ryan Cotterell, Reda Boumasmoud
Published: 8/26/2026, 4:00:00 AM
Categories: cs.FL, cs.CL, cs.LG
arXiv:2606.01765v3 Announce Type: replace-cross Abstract: What formal languages can a recurrent neural language model recognize? Formal results in the literature conflict: some authors report Turing-completeness, while others show equivalence to regular languages. The reason for this discrepancy is ...
261. Geometric bias in eigenspace perturbation under random heterogeneous noise ​
Author: Fengkai Liu, Ke Wang, Wanjie Wang
Published: 8/26/2026, 4:00:00 AM
Categories: math.ST, cs.LG, cs.NA, math.NA, math.PR, stat.TH
arXiv:2606.11263v2 Announce Type: replace-cross Abstract: Spectral methods rely on the stability of principal eigenspaces under random perturbations. Classically, this is quantified by the Davis-Kahan and Wedin theorems, which bound the eigenspace error via the operator norm of the noise and the rel...
262. Incremental Learning in Mirror Flows ​
Author: Rapha"el Berthier, Loucas Pillaud-Vivien
Published: 8/26/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ML
arXiv:2606.23198v2 Announce Type: replace-cross Abstract: We study mirror flows generated by a convex quadratic loss and a general convex lower semicontinuous mirror potential. We show that, when initialized near the boundary of the domain of the mirror potential, their rescaled trajectories converg...
263. Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop ​
Author: Chenmu Zhang, Boris I. Yakobson
Published: 8/26/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.AI, cs.LG
arXiv:2606.29717v3 Announce Type: replace-cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has produced standard public benchmarks and many published machine-learning models for the task (D...
264. Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages ​
Author: Lucas Pinto
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.06596v3 Announce Type: replace-cross Abstract: Trusted monitoring is a central defense in AI control: a cheaper trusted model scores an untrusted model's actions for sabotage, and the most suspicious are audited or deferred. Such monitors are evaluated against one or two untrusted models,...
265. Continual Learning With Participation Privacy: An Auditable Buffering-Aggregation Recipe ​
Author: T-H. Hubert Chan, Elaine Shi, Mengshi Zhao, Mingxun Zhou
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.07209v2 Announce Type: replace-cross Abstract: Modern federated and streaming learning systems often release intermediate models, so privacy must hold for the full trajectory under adaptive interaction. Motivated by participation privacy, we study single-edit neighboring user streams, whe...
266. Tool-Making and Self-Evolving LLM Agents in Low-Latency Systems ​
Author: Kalle Kujanp"a"a, Ning Liu, Shahnawaz Alam, Yeshwanth Reddy Sura, Tianyu Yang, Kristina Klinkner, Shervin Malmasi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SE
arXiv:2607.08010v2 Announce Type: replace-cross Abstract: Production LLM agents often waste latency and reliability by regenerating code for the same procedural steps on every request. We replace this inference-time coding loop with an agentic tool-making pipeline that compiles repeated SOP steps in...
267. Covariance-Boosted Gaussian Processes for Spatiotemporal Irregularities ​
Author: Jeremy Ovadia
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, physics.space-ph, stat.ME
arXiv:2607.23018v2 Announce Type: replace-cross Abstract: Nonstationary Gaussian process (GP) models are powerful tools for capturing input-dependent variability by adapting to observed data. However, with limited sampling and highly parameterized covariance structure, they are often prone to overfi...
268. Gated Recurrent Transformers: Expressive Depth through Recurrent Modulation in Transformers ​
Author: Amr Hegazy, Amr Alanwar, Mostafa Elhoushi
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.15062v3 Announce Type: replace-cross Abstract: Scaling transformer language models creates an inherent tension between expressivity and memory efficiency. While unique weights across layers preserve functional specialization---from input-grounding to abstract refinement---they incur a sub...
269. GOD: Enhancing Generalization via Deep Grafting for Sequential Recommendation ​
Author: WooJoo Kim, JunYoung Kim, JaeHyung Lim, HwanJo Yu
Published: 8/26/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.16073v2 Announce Type: replace-cross Abstract: Sequential recommenders often struggle with sparse and noisy histories, limiting generalization to unseen interactions. Knowledge distillation mitigates this by transferring dense supervision from a teacher to a student. However, most distill...
270. Representation Is Not Enough: Body-Localized Thermal Evidence for Contactless Stress and Craving Sensing in Opioid Use Disorder ​
Author: Sachin Deb, Harshit Sharma, Asif Salekin
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.16087v2 Announce Type: replace-cross Abstract: Removing wearables from physiological monitoring also removes their supervision: the signal indicating where and when a stress response occurred. Contactless stress sensing therefore becomes a weakly supervised evidence-localization problem, ...
271. What You Can't See Is What You Learn: Slot-Selective Evidence Masking Favors Compositional Generalization in Shared-Genome Language-Model Societies ​
Author: Narcis Marincat
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA
arXiv:2608.20054v3 Announce Type: replace-cross Abstract: Multi-module neural systems often expose every module to the full input. We test whether a slot-selective evidence-masking regime -- restricting each module to its own evidence span -- changes which solutions gradient-based training discovers...
272. If It Walks Like an Arbitrage: Protocol-Agnostic Detection with Decidable Structural Equivalence ​
Author: Adam Khayam, Hamid Kolli, Mohamed Iguernlala, \c{C}agdas Bozman
Published: 8/26/2026, 4:00:00 AM
Categories: q-fin.CP, cs.CR, cs.LG
arXiv:2608.20377v2 Announce Type: replace-cross Abstract: Whether a transaction performed an arbitrage, and by which route, is a question asked of its execution trace after the fact. We conjecture that such traces admit a normal form on which questions of this kind become queries, and we test it by ...
273. Gauss--Hermite Quadrature for Gaussian-Mixture Entropy with an Action-Space Hermite Surrogate ​
Author: Jae Wan Shim
Published: 8/26/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT
arXiv:2608.21467v2 Announce Type: replace-cross Abstract: Gaussian distributions are used to model uncertainty in signals and states, and Gaussian mixtures are often used when the underlying distribution is multimodal. Unlike a single Gaussian, a Gaussian mixture generally has no closed-form express...
274. Inferring Action from Future Latent State for Robotic Manipulation ​
Author: Fenghao Lei, Zhixiong Huang, Long Yang, Jiabao Chen, Peilin Huang, Han Fu, Zhuo Li, Xiaoxue Ren
Published: 8/26/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG
arXiv:2608.22067v2 Announce Type: replace-cross Abstract: World-Action Models (WAMs) build robot control on video-generation backbones, which jointly predict dense future visual trajectories and robot actions. We argue that video generation is an unnecessary intermediate objective for world-action m...
275. Autonomous Cyber Defense: Real-Time Attack Detection and Mitigation in Software-Defined Networks Using Machine Learning ​
Author: Alexandre Amaral, Fernando Moro, Ana Malheiro
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.22075v2 Announce Type: replace-cross Abstract: Adversaries now move faster than manual response processes can absorb. The average eCrime breakout time, that is, the interval between initial access and the first lateral movement to another host, fell to 29 minutes in 2025, a 65% increase i...
276. When More References Hurt: Contamination-Aware DINOv2 Memory Banks for Few-Shot Steel Defect Detection ​
Author: Hannaneh Kalantary, Javad Khoramdel
Published: 8/26/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.22082v2 Announce Type: replace-cross Abstract: Patch-memory anomaly detectors assume that their reference bank is normal, an assumption that is difficult to guarantee when additional industrial images are unverified. We study whether a few trusted normal images can safely recover useful n...
277. Apodex 1.1: Scaling Agentic Intelligence for Complex Work ​
Author: B. An, B. Li, B. Wang, B. Zhang, B. L. Wang, C. Feng, C. Wei, C. Xue, C. Zhang, D. Ng, D. Ye, E. Min, F. Chen, F. Liu, F. Yang, F. Ye, G. Sun, H. Ji, H. Xu, H. Yang, H. Ye, H. Zhang, H. Zhao, J. Li, J. Lin, J. Xia, K. Jin, K. Wang, K. Yang, L. Bing, L. Lei, L. Su, Le. Wang, Lu. Wang, N. Wang, Q. Ren, Q. Yang, R. Li, S. Bai, S. Du, S. Li, S. Lin, S. Nie, S. Wang, S. Zhang, S. Z. Wang, T. Ge, Ta. Q. Fang, Ti. Q. Fang, W. Fang, W. Li, W. Zhang, X. Chen, X. Li, X. Tang, X. Wang, X. Xu, X. Zhang, X. Q. Wang, X. Y. Wang, Y. Deng, Y. Gao, Y. Hu, Y. Li, Y. Sui, Y. Wang, Y. Xiao, Y. Zhang, Y. Zhou, Z. Chen, Z. Cheng, Z. Feng, Z. Liang, Z. Liu, Z. Zhang
Published: 8/26/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.23283v2 Announce Type: replace-cross Abstract: General-purpose language models can reason and synthesize knowledge, but complex work also requires sustained interaction with files, information sources, and executable code, together with state maintenance, failure recovery, and verifiable ...