arXiv cs.LG - 2026-08-13 ​
261 items collected.
1. FarSky: Task-Aware Latent-Space Coupling for Generative Intra-Hour Solar Forecasting ​
Author: Yann Fabel, Bijan Nouri, Milon Miah, Niklas Blum, Luis F. Zarzalejo, Julia Kowalski, Robert Pitz-Paal
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph, stat.AP
arXiv:2608.11254v1 Announce Type: new Abstract: Accurate solar irradiance forecasting is essential for the reliable integration of photovoltaic power into modern electricity grids. All-sky imagers (ASI) provide high-resolution observations of clouds, making them well suited for intra-hour forecastin...
2. Why AI Detection Fails for Academic Integrity ​
Author: Jonathan A. Karr Jr, Grigorii Khvatskii, Ting Hua, Nitesh V. Chawla
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CY
arXiv:2608.11256v1 Announce Type: new Abstract: Institutions use commercial AI detectors for academic integrity, yet detectors cannot distinguish AI editing from full LLM drafts and may treat both as misconduct. In a controlled study of published English abstracts (four domains; 2013 to 2015 vs. 202...
3. Basin: Efficient and Extensible Numerical Optimization in Rust ​
Author: Johan Larsson
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2608.11279v1 Announce Type: new Abstract: Basin is a numerical optimization library for the Rust programming language. Numerical optimization is the task of finding the inputs that minimize a function, and it is a fundamental element across the sciences: fitting a model to data, calibrating a ...
4. Federated Learning for Distributed CNC Tool Wear Prediction ​
Author: Afsana Khan, Morris Stallmann, Marcin Pietrasik, Charis Kouzinopoulos, Anna Wilbik
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11281v1 Announce Type: new Abstract: Tool wear prediction is an important task in CNC machining, where accurate monitoring of tool condition supports product quality and process reliability. Machine learning methods have shown potential for this task, but their use in industrial environme...
5. Terminal Symmetry as a Decision Resource: Statewise Refinement for Anytime Verified Construction ​
Author: Yi Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11318v1 Announce Type: new Abstract: Many sequential construction tasks exhibit exact symmetry at completion while their execution remains directed and history-dependent. We develop a decision-resource view of terminal symmetry: process evidence supplies directionality, terminal correspon...
6. Contextual Quality-Diversity Evolutionary Reinforcement Learning for HVAC Control in Tropical Commercial Buildings ​
Author: Tran Le Vu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE, math.OC
arXiv:2608.11324v1 Announce Type: new Abstract: This paper proposes a contextual quality-diversity evolutionary reinforcement-learning controller, CQD-ERL, for the supervisory control of a tropical, water-cooled chiller plant and its associated air side. Rather than converging to a single scalarised...
7. Long-Horizon Forecasting of Complete Financial Statements with Forma ​
Author: Travis L. Johnson, Jiannan Jiang, Soumyabrata Chaudhuri, Yihao Chen, Lauren Falvey, Donal O'Cofaigh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, q-fin.CP
arXiv:2608.11327v1 Announce Type: new Abstract: Specialist training beats generalist scale when forecasting financial statements. To our knowledge, no prior work jointly forecasts complete financial statements beyond one year, yet in a discounted-cash-flow valuation most firm value sits past that wi...
8. Weightless Fine-Tuning: Personalizing LLMs via Logit-Space Transport ​
Author: Bohan Zhang, Anqi Ni, Yixin Wang, Paramveer S. Dhillon
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.11342v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is a standard approach for adapting LLMs to a target distribution, but in settings such as personalization, where each author requires separate weight access, optimization, storage, and retraining, its costs become prohibit...
9. Dynamics Models for Offline Hyperparameter Selection in Real-World RL ​
Author: Jordan Coblin, Han Wang, Martha White, Adam White
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11349v1 Announce Type: new Abstract: A key obstacle to deploying reinforcement learning in real-world systems is hyperparameter selection, particularly when simulators are unavailable and online experimentation is costly. Prior work has proposed calibration models trained on offline data ...
10. Market-Information-Aware Gated-LoRA of Foundation Models for Transferable Day-Ahead Electricity Price Forecasting ​
Author: Hang Fan, Wei Wei, Shengwei Mei
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11359v1 Announce Type: new Abstract: Electricity price forecasting is crucial for market participants but remains difficult because prices are volatile, market-specific, and closely tied to anticipated system conditions. Existing supervised methods depend largely on market-specific histor...
11. Lifecycle-Optimal Tokenization: Vocabulary Size as a Deployment-Regime-Dependent Infrastructure Parameter ​
Author: Rima Mittal, Ankit Gubrani, Satyanarayana Kakollu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.PF
arXiv:2608.11361v1 Announce Type: new Abstract: Tokenizer vocabulary size is a foundational design choice in large language model (LLM) infrastructure, yet it is typically fixed at training time based on convention rather than deployment analysis. We show that the cost-optimal vocabulary is not a co...
12. PAIR: Pairwise-Aware Inclusion Reweighting for Adaptive Rollout Allocation in RLVR ​
Author: Pixel Nomand, Elena Voss, Marcus Hale, Sofia Reyes
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11368v2 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) spends most of its compute generating groups of long reasoning trajectories. Recent allocators reduce this cost by assigning budgets to prompts, rollouts, or tokens according to a pointwise notion o...
13. Towards an approach to multivariate outlier detection for District Heating System data ​
Author: Rajko Turudija, Du\v{s}an Stojiljkovi'c, Milan Zdravkovi'c, Marko Ignjatovi'c
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2608.11375v1 Announce Type: new Abstract: In this paper, we test different methods for multivariate detection of outliers in the data of transmitted heat energy in the selected substation of local District Heating System, by also considering outside ambient temperature, namely Z-score (univari...
14. Reoptimization Algorithms for Contextual Bandits with Knapsack Constraints ​
Author: Zhen Xu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2608.11383v1 Announce Type: new Abstract: We study new algorithms for Contextual Bandits with Knapsack. In these problems, there are finitely many types of customers, products, and resources. Each product is made from a fixed combination of resources, and resources have finite capacity. A deci...
15. Mechanism Design for Generative Engines: From Exploitation toward Win-Win Outcomes ​
Author: Chen Xu, Zitian Guo, Chenyan Xiong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11390v1 Announce Type: new Abstract: Generative engines are reshaping the web ecosystem by making citations a key mechanism for allocating attention, attribution, and downstream value. This creates a strategic tension: content providers are incentivized to optimize for model citation, whi...
16. Unmasking Toxic Mimicry in Medical Offline Reinforcement Learning for ICU Sepsis Management via Counterfactual Clinical Audits ​
Author: Hangqi Ren, Junyi Liao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CY
arXiv:2608.11410v1 Announce Type: new Abstract: Offline reinforcement learning (RL) offers considerable promise for optimizing ICU treatment decisions, yet standard evaluation metrics Mean Squared Error (MSE) and Fitted Q-Evaluation (FQE) assess only behavioral imitation and cannot detect Toxic Mimi...
17. Diffusion-Based Data-Driven Assortment Optimization ​
Author: Junyi Liao, Xiaohui Jiang, Zhengwei Tong, Ethan X. Fang, Vahid Tarokh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11419v1 Announce Type: new Abstract: Assortment optimization is a fundamental problem in revenue management, typically addressed using parametric choice models such as the multinomial logit (MNL) and its variants. While these models enable tractable formulations, their performance is sens...
18. Analysis of Federated Aggregation under Model Poisoning and Backdoor Attacks: A Reconstructed Cross-Dataset and Cross-Architecture Benchmark ​
Author: Soumya Mazumdar, Vineet Kumar Rakesh, Tapas Samanta
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.CV
arXiv:2608.11423v1 Announce Type: new Abstract: Robust comparisons of federated aggregation methods require joint consideration of predictive performance, threat definitions, metric semantics, and execution provenance. A 500-cell seed-1 evaluation matrix was reconstructed across five aggregation met...
19. Click2Poly: A VLM for vector mapping buildings and walls ​
Author: Nicolas Girard, Jawher Ben Abdallah, Arno Gobbin, Liuyun Duan, Sacha Lepretre
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.11424v1 Announce Type: new Abstract: Accurate vector mapping of buildings and walls is critical for geospatial applications but remains a labor-intensive process. While recent deep learning methods have improved automatic extraction, in order to meet cartographic standards they always req...
20. Three Tokens Force Exponential Feature Rank in Nonnegative Kernel Attention ​
Author: Vicente Opazo
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11427v1 Announce Type: new Abstract: Full attention exposes every token pair, whereas kernel attention compresses a sequence into a fixed-dimensional sketch. We show that this distinction becomes exponential at the first context length containing two competing candidates. On Min-IP over B...
21. AutoGrable: What Is a Good Graph for a Table? ​
Author: Tamara Cucumides, Floris Geerts
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11431v1 Announce Type: new Abstract: Graph learning presupposes a graph, and tables and relational databases do not come with one. Applying a GNN to them requires deciding which entities become nodes, which of them to connect, and through which relations---a decision made by hand, by sche...
22. Variational Parameter Calibration with Physics-Aware Latent-Space Surrogates ​
Author: Qiyao Zhou, Xujia Zhu, Pierre Joli, Yu Cong, Sibo Cheng
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph, physics.data-an, physics.flu-dyn
arXiv:2608.11435v1 Announce Type: new Abstract: Forward and inverse modeling of parametric dynamical systems requires surrogate models that are not only accurate for state prediction, but also informative for parameter calibration. However, a systematic end-to-end differentiable formulation for coup...
23. XGBoost "is all you need": the case of forecasting transmitted heat energy in District Heating Systems ​
Author: Milan Zdravkovi'c
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2608.11446v1 Announce Type: new Abstract: This paper presents a comparative study of two distinct approaches, XGBoost and Long-Short Term Memory (LSTM), for forecasting transmitted heat energy in District Heating Systems (DHS). The objective is to explore scenarios in which conventional ML alg...
24. PAC-Bayes Beyond Parameter Space: Behavioral Equivalence, Z-Information, and Exact Complexity Decomposition ​
Author: Vasant G. Honavar, Satish Kumar Keshri, Neil Ashtekar, Zehao Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11465v1 Announce Type: new Abstract: PAC-Bayes theory provides generalization guarantees by controlling the Kullback--Leibler (KL) divergence between posterior and prior distributions over a chosen hypothesis representation. However, predictive risk depends only on the predictive behavior...
25. Dual-Primal Graph VAEs for Noisy Label Aggregation ​
Author: Patrick Stinson, Nikolaus Kriegeskorte
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11473v1 Announce Type: new Abstract: Inferring the ground-truth from noisy crowdsourced labels is an important theoretical and practical problem. Neural network-based methods offer an alternative to classical Bayesian models which require specifying a family of generative models used for ...
26. Convergence Guarantees of Gradient Descent for Neural Networks via Generalized Lipschitz Smoothness ​
Author: Siqiao Mu, Diego Klabjan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11479v1 Announce Type: new Abstract: We establish convergence guarantees of gradient descent for general feedforward neural networks of arbitrary width or depth, with no special requirements on the initialization or dataset. We only assume that the activation functions are Lipschitz smoot...
27. Defending against Model Extraction for GNNs with Model Reprogramming ​
Author: Yan Wen, Zhenyi Wang, Heng Huang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.IR
arXiv:2608.11495v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) serve as the backbone for high-stakes applications in Machine-Learning-as-a-Service (MLaaS). Still, their black-box deployment exposes them to Model Extraction (ME) attacks, in which adversaries steal intellectual property ...
28. HyperFix: Combinatorial Nonlinear Correction for Task Vector Merging ​
Author: Hyo Seo Kim, Ren Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11499v1 Announce Type: new Abstract: Task vectors enable model merging without joint retraining. In practice, the subset of task vectors to be merged may vary, but many existing methods use scalar tuning for a particular subset, requiring repeated tuning across subsets and restricting tas...
29. RelShap: Relationally Consistent Shapley Explanations ​
Author: Seungeun Lee, Joao Fonseca, Julia Stoyanovich
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11508v1 Announce Type: new Abstract: Machine learning pipelines commonly flatten relational data into single-table representations, discarding structural constraints. Widely used Shapley value-based feature attributions then rely on feature independence, evaluating the model on combinatio...
30. Let it Cook: Learning to Wait in Sequential Decision Making ​
Author: Christopher Watson, Arjun Krishna, Dinesh Jayaraman, Rajeev Alur
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11511v1 Announce Type: new Abstract: In sequential decision making, an agent typically observes its environment and acts at every timestep. However, such active participation may not always be necessary; tasks such as brewing coffee include periods that are served equally well by letting ...
31. FLARE++: Low-rank attention with dynamic attention routing ​
Author: Vedant Puri, Yongjie Jessica Zhang, Levent Burak Kara
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11519v1 Announce Type: new Abstract: Full self-attention is a strong token mixer for PDE surrogates on irregular domains, but its quadratic cost limits its use on high-resolution problems. Efficient latent-attention models such as the Fast Low-rank Attention Routing Engine (FLARE) avoid t...
32. Hierarchical Federated Transfer Learning in Digital Twin-Based Vehicular Networks ​
Author: Qasim Zia, Saide Zhu, Haoxin Wang, Zafar Iqbal, Yingshu Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11532v1 Announce Type: new Abstract: In recent research on the Digital Twin-based Vehicular Ad hoc Network(DT-VANET), Federated Learning (FL) has shown its ability to provide data privacy. However, Federated learning struggles to adequately train a global model when confronted with data h...
33. Robust Ambiguity Detection (RAD) From Model- and Feature-Space Consistency ​
Author: Manya Singh, Mark T. Keane, Arjun Pakrashi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11541v1 Announce Type: new Abstract: Machine learning models should be robust, in the sense of remaining predictively consistent under permissible variations. A model's predictions should ideally remain unchanged when it is replaced by a functionally equivalent one, or when its inputs are...
34. Certifying What Helps Customer-Return Timing: A Screen-and-Confirm Test for Conditioning Signals, and Why Decay Is Nearly Enough ​
Author: Sang Su Lee, Vineeth Loganathan, Shishir Dash, Vijay Raghavan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.AP
arXiv:2608.11555v1 Announce Type: new Abstract: Practitioners enrich customer-return models with ever more signals (lifetime value, category, recency/frequency, calendar, geography), and the temporal-point-process (TPP) literature follows suit with covariate- and external-covariate-conditioned inten...
35. When Offline Evaluation Misleads: A Diagnostic Protocol for Reward and Policy Selection in Delayed-Feedback Contextual Bandits ​
Author: Sang Su Lee, Vineeth Loganathan, Shishir Dash, Vijay Raghavan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11560v1 Announce Type: new Abstract: Personalizing marketing messages with contextual multi-armed bandits (CMABs) drives real business value, yet the objective that ultimately matters - a downstream conversion - is observed only weeks later, too late to drive online learning. Teams theref...
36. Sparse and robust geometric twin support vector machine via asymmetric RoBoSS loss function ​
Author: Kai Qi, Xinji Huang, Hongchun Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11567v1 Announce Type: new Abstract: In real-world scenarios, the training data usually contains redundant features, label noise and feature noise, which provide severe challenges for the efficiency of machine learning methods. Since standard support vector machine (SVM) adopts $l_2$-norm...
37. RECAST: A Machine-Learning Framework for Correction and Super-Resolution of Coarse-Grid PDE Solvers ​
Author: Maryam Reza, Farbod Faraji
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2608.11572v1 Announce Type: new Abstract: Coarse-grid numerical solvers can substantially reduce the computational cost of time-dependent PDE simulation, but under-resolution often degrades both the trajectory and the spatial fidelity of the solution. We introduce RECAST (Recurrent Error Corre...
38. Dion3: Full-Stack Orthogonal Updates ​
Author: Noah Amsel, Jack Zhang, Kwangjun Ahn, Ali Naeimi, Austin Feng, Berlin Chen, Tri Dao, John Langford
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11612v1 Announce Type: new Abstract: The Muon optimizer incurs a significant overhead cost due to its cubic-time Newton-Schulz orthogonalization step. When weights are sharded, communication overhead compounds this computational cost, eroding the benefits of Muon in many settings. We pres...
39. A Local Sinkhorn Framework for Conditional Distribution Reconstruction of Multidimensional Random Fields ​
Author: Mingtao Xia, Qijing Shen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11613v1 Announce Type: new Abstract: In this paper, we propose a local Sinkhorn divergence framework for conditional distribution reconstruction of multidimensional random fields. By utilizing the debiased Sinkhorn divergence, our proposed approach develops a differentiable and computatio...
40. FM-LLM: A frequency-enhanced mixture-of-experts framework for adapting LLMs to time series forecasting ​
Author: Rentao Gu, Yihang Ding, Junjie Li, Yi Ding, Weijing Sang, Xiaoli Huo, Xin Qin, Yuefeng Ji
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NI, eess.SP
arXiv:2608.11623v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have spurred cross-modal solutions for time-series forecasting. However, existing methods rely heavily on textual prompts for modality alignment-introducing nontrivial computational overhead and failing t...
41. Transferable Above-Ground Biomass (AGB) Estimation Model from Multi-Sensor Data with Sparse Field Calibration ​
Author: Pann Thinzar Seint, Bryan Atwood, Subas Chhatkuli
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.11638v1 Announce Type: new Abstract: Spatially continuous quantification of forest above-ground biomass (AGB) is what makes carbon accounting credible and mitigation strategies actionable. While field inventories provide high localized accuracy, they are spatially sparse; conversely, spac...
42. Towards a Formal Definition of Agent Memory: Basis, Span, Optimality, and the Sequential Memory Problem ​
Author: Hongyao Tang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11654v1 Announce Type: new Abstract: Despite the wide deployment of memory in large-model agents, there is no unified formal account of what a memory is or when it is optimal. This paper takes a first step toward this account. The central idea is that memory is a basis, knowledge is its s...
43. Continuous-Latent Predictive Modeling with Semantic Alignment for EEG-Language Foundation Models ​
Author: Myeong-Ju Cho, Hye-Bin Shin, Seo-Hyun Lee, Seong-Whan Lee
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11656v1 Announce Type: new Abstract: Recent advances in EEG foundation models have demonstrated the potential of large-scale pretraining to enable generalizable neural decoding across subjects, recording environments, and datasets. However, dominant pretraining paradigms face key challeng...
44. Is Per-Agent Policy Composition Safe? Rethinking Successor-Feature Transfer in Cooperative Multi-Agent Reinforcement Learning ​
Author: Zijian Zhao, Sen Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA
arXiv:2608.11658v1 Announce Type: new Abstract: Many reinforcement learning systems, from fleet management to traffic signal control, must serve an objective that changes dynamically after deployment, and retraining a policy for each new objective is prohibitively expensive. For a single agent, this...
45. Low-Interaction-Rank Learning: Unifying Multiplicative Dual-Encoder Heads ​
Author: Zijian Zhao, Sen Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11661v1 Announce Type: new Abstract: A multiplicative dual-encoder network computes a real-valued output for a pair of inputs as the inner product of their separate encodings. This architecture has been developed independently in operator learning, bipartite matching, contrastive vision-l...
46. Rubric Dropout: A Simple Way to Mitigate Reward Hacking in Rubric-as-Reward RL ​
Author: Minglai Yang, Xinyu Guo, Utkarsh Tyagi, Mian Zhang, Razvan Dumitru, Sunjie Hou, Yunzhong He, Daniel Yue Zhang, Ying Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.11669v1 Announce Type: new Abstract: Reinforcement learning against rubrics, lists of criteria graded by an LLM judge, has become a standard way to post-train language models on tasks with no deterministic answer. The rubric, however, is a fixed proxy for quality, never a complete descrip...
47. GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs ​
Author: Kai Yang, Jingwei Xu, Wanyu Wang, Kai-Yuan Guo, Zhenbo Yu, Yi Wang, Yu Qiao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11674v1 Announce Type: new Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they frequently suffer from training instabilities, cross-task capability degradation, and response-length inflation. Although prior work has characterize...
48. FunnelCausalNet: Funnel-aware Joint Conversion-Revenue Uplift for Multi-tier Coupon Allocation ​
Author: Yu Zhang (AMap Alibaba Group, Beijing, China), Zhihan Wang (AMap Alibaba Group, Beijing, China), Guanlin Chen (AMap Alibaba Group, Beijing, China), Min Jiang (AMap Alibaba Group, Beijing, China), Shuai Li (AMap Alibaba Group, Beijing, China)
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2608.11675v1 Announce Type: new Abstract: Coupon campaigns seek to lift both conversion and revenue, but gross merchandise value (GMV) follows a deterministic funnel from conversion to conditional order value and is zero-inflated and heavy-tailed. We propose FunnelCausalNet, an uplift estimato...
49. Drift and Dependence: Layer-wise Information-Theoretic Bounds for Replay-Based Continual Learning ​
Author: Tieliang Gong, Zhongbo Zhang, Wen Wen, Yong-Jin Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.11690v1 Announce Type: new Abstract: Continual learning must absorb new tasks without erasing old ones, and replay---mixing a small buffer of past examples into current training---is among the most effective remedies for catastrophic forgetting. Yet its generalization behavior is shaped b...
50. LEMUR: Latent Entropy-aware Multimodal Unlearning via Visual-anchored Reasoning Redirection ​
Author: Xinhao Zhong, Yuxia Qiao, Junhao Li, Hao Fang, Yi Sun, Bin Chen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.11691v1 Announce Type: new Abstract: Reinforcement-learning (RL) post-training equips multimodal large reasoning models (MLRMs) with exploratory chains of thought (CoT), substantially improving visual reasoning. However, we find that this capability introduces a distinct privacy vulnerabi...
51. REOPD: Reliability-Adaptive Reward Extrapolation for On-Policy Distillation ​
Author: Yang Sun, Lichao Ma, Houyuan Qin, Yuxin Liu, Hanyang Lu, Yao Zhu, Pinlong Cai, Guohang Yan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11698v2 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own trajectories under dense token-level supervision from a teacher. Reward-extrapolation methods such as ExOPD amplify the teacher-reference log-likelihood ratio to move beyond direct imitation, but...
52. Consolidator: Learning Persistent Routed Memory Across Context Boundaries ​
Author: Sungwoo Goo, Hwi-yeol Yun, Sangkeun Jung
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11701v1 Announce Type: new Abstract: Copying short-term memory (STM) into a slower store can preserve state across a context boundary, but persistence alone does not ensure that the retained state influences subsequent memory access. We test this distinction in a Phasor Memory Network (PM...
53. Robust and Efficient Noisy-Label Time-Series Classification via Dynamic Time Warping Based Granular Ball Computing ​
Author: Ziqiang Li, Yun Liu, Gouhei Tanaka
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11704v1 Announce Type: new Abstract: Dynamic Time Warping (DTW)-based Nearest-Neighbor (NN) classifiers are effective for time-series classification but are vulnerable to mislabeled training samples and require numerous DTW computations during inference. We propose DTW-based Granular Ball...
54. High-dimensional Multi-objective Bayesian Optimization with Learned Variable Interactions ​
Author: Hongyan Wang, Jiayu Huang, Haotian Zheng, Xin Gao, Chi Ding, Ying Liu, Xia Wang, Qing Xu, Keqiang Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11713v1 Announce Type: new Abstract: Multi-objective Bayesian optimization (MOBO) is effective in identifying the Pareto fronts for expensive black-box problems. However, most current MOBO approaches are limited to low-dimensional decision space due to its exponential sampling complexity....
55. Chain-of-Thought Shows the Path to a Tree: Realizing Branching Complexity ​
Author: Debanjan Dutta, Anish Chakrabarty, Swagatam Das
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11716v1 Announce Type: new Abstract: Chain of Thought (CoT) lifts the expressive ceiling of bounded-depth Transformers, with characterizations tying the number of CoT steps to circuit complexity classes. What remains largely missing are concrete instantiations with explicit, depth-bounded...
56. Epiplexity Guided Data Selection and Generation for Out-of-Distribution Generalization ​
Author: Ellen Su, Andres Potapczynski, Shikai Qiu, Edward Hughes, Andrew Gordon Wilson
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.11746v1 Announce Type: new Abstract: Modern systems are increasingly expected to transfer across tasks not specified during training. What data facilitates generalization in these new, unanticipated settings? One hypothesis is that data with more structural information could contain share...
57. MOON: Multi-Objective OrthoNormalized Updates for Multitask Learning ​
Author: Shiji Zhou, Kunlin Lyu, Lei Zhang, Ruodong Wang, Yifan Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.11749v1 Announce Type: new Abstract: Multi-objective optimization (MOO) has demonstrated significant success in multi-task learning by mitigating task conflicts through gradient manipulation. However, most existing methods flatten model parameters into vectors and perform gradient manipul...
58. TradingMoE: Routing the Right Experts in Evolving Markets ​
Author: Chang Zhou, Xingtong Yu, Minbin Huang, Zhennan Wu, Yuan Fang, Hong Cheng, Xinming Zhang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11785v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong potential for financial analysis and trading, but direct trading remains challenging because the predictive capabilities required can vary across assets, decision fields, and market conditions. Existing LL...
59. High-Order Liquid Evidence Encoding for Gradual GNSS Spoofing Detection in Autonomous Driving ​
Author: Muhammad Ayub Sabir, Junbiao Pang, Fatima Ashraf
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11790v1 Announce Type: new Abstract: Accurate Global Navigation Satellite System (GNSS)-based localization is essential for safe and reliable autonomous driving. However, spoofing attacks can manipulate vehicle position estimates. Continuous and subtle attacks are particularly difficult t...
60. Orientation, not magnitude: the causal structure of task-vector interference in merged language models ​
Author: Chencheng Zhu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11797v1 Announce Type: new Abstract: Model merging by task arithmetic works until it doesn't, and the field diagnoses why with magnitudes: layerwise representation bias, deviations from cross-task linearity, parameter overlap. Tracking the exact layerwise cross-term of merged LLMs through...
61. JAPE: Joint Anomaly Prediction and Intrinsic Explanation in Multivariate Time Series ​
Author: Yian Wei, Yuanyuan Yao, Lu Chen, Xiangmin Zhou, Tianyi Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11801v1 Announce Type: new Abstract: Multivariate time-series anomaly prediction aims to identify whether and when anomalies will occur over a future horizon from historical observations. Existing methods primarily characterize anomalies as deviations in future numerical values, which may...
62. Learning with Bilevel-Minimax Optimization for Efficient and Reliable Transfer Attacks ​
Author: Yaohua Liu, Yifan Guo, Jiaxin Gao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2608.11815v1 Announce Type: new Abstract: Transfer-based adversarial attacks craft adversarial examples using surrogate models to mislead black-box victim models. Beyond perturbation generation, transferability is fundamentally governed by the coupling of initialization, surrogate adaptation, ...
63. Towards Understanding On-Policy Distillation through the Lens of Test-Time Scaling ​
Author: Xinmu Ge, Zizhuo Zhang, Yu Huang, Jianing Zhu, Lin Yuan, Wanli Gu, Weichang Wu, Weiran Huang, Xiaolu Zhang, Bo Han, Jun Zhou, Jiangchao Yao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2608.11829v1 Announce Type: new Abstract: On-policy distillation (OPD) has emerged as a promising post-training technique for enhancing LLM reasoning. It is commonly believed to enable the student model to distill knowledge from a stronger teacher model, thereby expanding capabilities beyond t...
64. Kernel Methods for Learning Operators with Multiple Inputs and Outputs ​
Author: Adrien Weihs, Chunyang Liao, Jingmin Sun, Hayden Schaeffer
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH
arXiv:2608.11831v1 Announce Type: new Abstract: Learning mappings between infinite-dimensional objects is a central challenge in scientific machine learning. We introduce a general kernel-based encoder-decoder framework for operator learning that separates observation, representation, learning, and ...
65. Air Quality Station Simulation via LSTM and Attention-Based Modelling ​
Author: Alexander Kostadinov, Petar O. Hristov, Dessislava Petrova-Antonova
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11839v1 Announce Type: new Abstract: Poor air quality in urban areas is driven by a complex chain of processes and presents a significant public health concern. To better understand and control the mechanisms that determine air quality, cities deploy networks of measurement stations, and ...
66. Small-Scale Experiments: Are We There Yet? ​
Author: Nicholas Lourie, Kyunghyun Cho, Karen Ullrich, Sanae Lotfi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11859v1 Announce Type: new Abstract: Scaling laws promised cost-effective experiments; six years later, they have yet to fully deliver. Instead, researchers have found them unreliable at small scales (starting at 4M parameters) and concluded that sizable models cannot be avoided. We show ...
67. Forward and Inverse Virtual Metrology for Phototransistor Gain: A Hierarchical, Uncertainty-Aware Approach for Small Production Datasets ​
Author: Mahshid Amirabgir, Lorenza Ferrario, Paolo Conci, Mahdieh Amirabgir, Giancarlo Orengo
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.SY, eess.SY
arXiv:2608.11868v1 Announce Type: new Abstract: The customization, optimization and stabilization of the process flow of a silicon bipolar phototransistor commits months of cleanroom time before a finished device can be measured, so a model that predicts device gain from process parameters before a ...
68. DCM Bandits: Multiplayer Information Asymmetric Cascading Bandits for Multiple Clicks ​
Author: Andy Wang, Charlton Shih, William Chang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11873v1 Announce Type: new Abstract: In this work, we extend the Dependent Click Model (DCM) Bandits to a multiplayer information-asymmetric setting, where multiple agents interact with a shared ranked list and may observe multiple clicks per session, introducing new challenges for select...
69. Disentangling the Expressivity of RoPE ​
Author: Selim Jerad, Anej Svete, Jiaoda Li, Ryan Cotterell
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.FL
arXiv:2608.11909v1 Announce Type: new Abstract: Two accounts recur in explanations of the success of rotary position embeddings (RoPE). Expressivity studies associate periodic position information with modular predicates, whereas mechanistic and long-context studies emphasize positional anchors and ...
70. A Factor Graph Approach to Scalable Multi-Output Gaussian Process Regression ​
Author: Wouter W. L. Nuijten, Esther G. van Pelt, Albert Podusenko, .Ismail \c{S}en"oz, Wouter M. Kouw
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11917v1 Announce Type: new Abstract: Multi-output Gaussian process regression scales cubically in the number of observations times outputs, and dense kernel-matrix methods need bespoke handling whenever different outputs are observed at different inputs. We express multi-output Gaussian p...
71. Distillation of Foundation Models for Time-dependent PDEs ​
Author: Daniel Musekamp, Boshra Ariguib, Andrei Manolache, Mathias Niepert
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11937v1 Announce Type: new Abstract: Foundation models for time-dependent partial differential equations (PDEs) are trained on large and diverse collections of physical systems and can generalize effectively to new downstream tasks. After fine-tuning on only a few trajectories from a targ...
72. TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement ​
Author: Karim Aly, Alexei Sharpanskykh, Jacco Hoekstra
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11951v1 Announce Type: new Abstract: Extreme events in air transport, such as severe arrival delays and abnormal air times, cause cascading network disruptions with substantial operational, economic, and safety costs. Such events are rare in historical records, leaving insufficient traini...
73. LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation ​
Author: Zhixin Zhang, Xinke Jiang, Zhibang Yang, Weixuan Xu, Guohong Qiu, Xu Chu, Junfeng Zhao, Yasha Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.11967v1 Announce Type: new Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical capability in such settings is reflection: assessing trajectory progress, identifying missing evidence a...
74. TESLA: Taylor Expansion of Sinusoidal Learnable Activations ​
Author: Daehwa Ko, Jaehyeon Kim, Seunghyun Ham, Jay Hoon Jung
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.11970v1 Announce Type: new Abstract: The parity problem--deciding whether the number of ones in a binary vector is odd or even--remains challenging for standard neural networks due to linear inseparability and the need for global interactions. We propose TESLA, an activation defined as a ...
75. Remote Sensing and Machine Learning-Based Analysis of Land Use and Vegetation Change in Dhaka District, Bangladesh ​
Author: Muhammad Masud Tarek, Md. Alamgir Hossain, Md. Samiul Islam, Muntasir Hasan Kanchan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.12001v1 Announce Type: new Abstract: Rapid urbanization in Dhaka District, Bangladesh has triggered substantial alterations in land use and environmental conditions, necessitating systematic monitoring for informed urban planning and ecological sustainability. This study employs remote se...
76. Dual-Model Sentiment Analysis of Consumer Reviews in the Retail Coffee Sector Using Machine Learning and Deep Learning Approaches ​
Author: Muntasir Hasan Kanchan, Md. Alamgir Hossain, Md. Samiul Islam, Muhammad Masud Tarek
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2608.12007v1 Announce Type: new Abstract: Consumer reviews play an important role in shaping brand perception and business strategies, particularly in service-driven industries such as retail coffee. This study presents a comparative sentiment analysis framework for Starbucks customer reviews ...
77. Reducing Symmetry Increase in Equivariant Neural Networks ​
Author: Ning Lin, Jiacheng Cen, Anyi Li, Wenbing Huang, Hao Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12010v1 Announce Type: new Abstract: Equivariant Neural Networks (ENNs) have empowered numerous applications in scientific fields. Despite their remarkable capacity for representing geometric structures, ENNs suffer from degraded expressivity when processing symmetric inputs: the output r...
78. SoftWater: Class-Aware Rate Allocation for Softmax Quantization ​
Author: Joao V. Cavalcanti, Ashia C. Wilson
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12026v1 Announce Type: new Abstract: Post-training quantization pipelines routinely leave the softmax output layer in high precision. Yet in small LLMs with modern vocabularies, the head holds 15--30% of all parameters, so a nominal ``2-bit'' model with an fp16 head can store several tim...
79. Uncertainty-Aware Probabilistic Constrained Clustering from Entangled Pairwise Supervision ​
Author: Shaojie Zhang, Ke Chen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.12027v1 Announce Type: new Abstract: Pairwise constrained clustering typically relies on hard must-link/cannot-link labels, whereas realistic pairwise supervision may be real-valued and entangle intrinsic ambiguity, expert judgment, and stochastic corruption. Existing deep constrained clu...
80. Clustered Randomized Smoothing for Stochastic Prediction Functions ​
Author: Eduardo Figueiredo, Frederik Mathiesen, Julian Schumann, Jens Kober, Arkady Zgonnikov, Luca Laurenti
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2608.12037v1 Announce Type: new Abstract: Modern stochastic predictors can model rich, multi-modal outcome distributions. However, this expressive power comes with challenges in ensuring robust predictions $-$ a critical requirement in safety-critical domains. Randomized smoothing is a leading...
81. Towards Truly Unsupervised Evaluation of Feature Selection ​
Author: Hafiz Saud Arshad, Muhammad Rajabinasab, Arthur Zimek
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12057v1 Announce Type: new Abstract: Feature selection is one of the most important and fundamental tasks in data mining, tackled by a family of methods with an established set of evaluation techniques to measure the quality of a specific method. Most of the methods commonly used for the ...
82. Faithful, Sufficient and Understandable: Rethinking Graph Counterfactual Explanations via Discrete Diffusion Inversion ​
Author: David Bechtoldt, Sidney Bender
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.12083v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve strong predictive performance on graph-structured data across domains such as chemistry, biology, and network analysis, yet they provide no intrinsic explanation of their predictions. This limits their adoption in h...
83. NAE: Normalizing AutoEncoder ​
Author: Muhammad Abdur Rafae, Niels Landwehr
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12084v1 Announce Type: new Abstract: We consider the setting of Normalizing flows with approximate inverses, an established paradigm spanning both full-dimensional ($d=D$) and bottleneck ($d<D$) settings, and group these models under the term flow autoencoders. We present a theoretical in...
84. Task- and dataset-specific information in protein language models ​
Author: Roman Joeres, Ilya Senatorov, Anastasia Kolchina, Dietrich Klakow, Olga V. Kalinina
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM
arXiv:2608.12090v2 Announce Type: new Abstract: Protein language models (PLMs) have transferred the latest advances from natural language processing to computational biology. These models, trained on large corpora of protein sequence data, are widely used to translate amino acid sequences into laten...
85. Confidence Calibration of Deep Learning Systems ​
Author: Coby Penso
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2608.12100v1 Announce Type: new Abstract: In high-stakes applications, reliable confidence estimates are as important as the predictions themselves. Confidence calibration ensures that predicted probabilities reflect the likelihood of correctness, making it essential for safe deployment of dee...
86. Beyond Parameter Space: NTK-Guided Personalized Aggregation for Robust Federated Learning ​
Author: Mirko Konstantin, Stefan Zachow, Anirban Mukhopadhyay
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12108v2 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed clients while keeping data local. A central challenge is determining which client updates are beneficial for aggregation with respect to each client's target domain. Existi...
87. Attractor Image-Based Deep Learning of Arterial Pulse Waves for Age Classification ​
Author: Sara Vardanega, Patrick Segers, Philip Aston, Ernst Rietzschel, Jordi Alastruey, Manasi Nandi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12117v1 Announce Type: new Abstract: Arterial pulse waveform morphology evolves with age, reflecting structural and functional changes in the cardiovascular system. Thus, vascular age is a valuable surrogate marker of cardiovascular health, and premature vascular ageing can indicate incre...
88. Adversarial Resilience of Poisson-Process Submodular Maximization over Matroids: From Robust Offline Optimization to Full-Bandit Learning ​
Author: Vaneet Aggarwal
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CC, math.OC
arXiv:2608.12134v1 Announce Type: new Abstract: We study nonnegative submodular maximization subject to a general matroid when the offline algorithm is given an arbitrary controlled value oracle. Our main result is an adversarial resilience theorem for the Spiteful Greedy Swap Poisson Process (SGS-P...
89. HYDRA: Hyperbolic Dynamic Representation Architecture for Kolmogorov-Arnold Networks ​
Author: Zhao Su, Yuxin Xia, Haoran Li, Jun Shen, Qi Zhu, Qingguo Zhou, Binbin Yong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.12194v1 Announce Type: new Abstract: Kolmogorov-Arnold Networks (KANs) enhance nonlinear function approximation by replacing scalar weights with learnable univariate functions. However, assigning an independent function to every connection results in substantial parameter redundancy, limi...
90. ScreenShot: A Foundation Model for Few-Shot Combination Drug Screening ​
Author: Antoine de Mathelin, Christopher Tosh, Wesley Tansey
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12219v1 Announce Type: new Abstract: Treating patients with combinations of drugs reduces the risk of resistance to any individual drug. Finding effective combinations is difficult because the large search space makes combinatorial screens prohibitively expensive, time consuming, and ofte...
91. An Efficient Near-Optimal Algorithm for Adversarial $m$-Set Bandits ​
Author: Francesco Bacchiocchi, Tommaso Cesari, Roberto Colomboni
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12231v1 Announce Type: new Abstract: We study adversarial combinatorial bandits with $m$-set actions, where at each round the learner selects $m$ out of $d$ items and observes only the aggregate loss of the selected items. The resulting action set contains $K=\binom{d}{m}$ elements and ca...
92. Calibration Bets on the Past: Post-Training Quantization for Financial Time-Series Forecasting ​
Author: Junyi Ye, Ivy Gateri Wanjiku
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, q-fin.ST
arXiv:2608.12259v1 Announce Type: new Abstract: Financial forecasting models are typically developed in full precision, yet production deployment often requires low-precision inference to reduce memory and computational cost. Post-training quantization (PTQ) enables such deployment without retrainin...
93. Earth observation embeddings are effective sub-grid descriptors for probabilistic weather downscaling ​
Author: Pedro Sousa (Department of Computer Science, University of Cambridge), Will Tebbutt (Department of Engineering, University of Cambridge), Sadiq Jaffer (Department of Computer Science, University of Cambridge), Robin Young (Department of Computer Science, University of Cambridge), Anil Madhavapeddy (Department of Computer Science, University of Cambridge), Richard E. Turner (Department of Engineering, University of Cambridge)
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph
arXiv:2608.12271v1 Announce Type: new Abstract: Global weather reanalyses and forecasts resolve the evolving atmospheric state on coarse grids, but site-specific applications require predictions at arbitrary locations where near-surface conditions also depend on unresolved terrain and land-surface p...
94. A Framework for Designing Reward Functions: From Objectives to Features to Human-Aligned Reward Functions ​
Author: Di Yang Shi, W. Bradley Knox
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.12302v1 Announce Type: new Abstract: We present a formal process to enable non-experts to instantiate and iterate on human-aligned reward functions, i.e. reward functions that adhere to a given preference ordering over trajectories. Given a task described in natural language, our process ...
95. Redistribution-based Cost Inference Improves Sparse Safe Offline RL ​
Author: Ebenezer Gelo (University of the Witwatersrand), Geraud Nangue Tasse (University of the Witwatersrand), Steven James (University of the Witwatersrand), Benjamin Rosman (University of the Witwatersrand)
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.12306v1 Announce Type: new Abstract: Safe offline RL typically assumes access to dense per-step cost annotations, but in practice supervisors provide only trajectory-level stop-feedback: a binary signal at the first unsafe transition, with no per-step attribution. We frame this as a tempo...
96. AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses ​
Author: Cheng Qian, Wenting Zhao, Liangwei Yang, Heng Wang, Jielin Qiu, Heng Ji, Silvio Savarese, Huan Wang, Shelby Heinecke
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2608.12307v1 Announce Type: new Abstract: Recent work on distillation transfers the capabilities of large models to smaller ones often by updating the latter's parameters, through teacher forcing, on-policy distillation, and related training-time methods. In this paper, we ask whether such tra...
97. WavePhaseNet: A DFT-Based Method for Constructing Semantic Conceptual Hierarchy Structures (SCHS) ​
Author: Kiyotaka Kasubuchi, Kazuo Fukiya
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2602.14419v2 Announce Type: cross Abstract: This paper reformulates Transformer/Attention mechanisms in Large Language Models (LLMs) through measure theory and frequency analysis, theoretically demonstrating that hallucination is an inevitable structural limitation. The embedding space functio...
98. Reproducing and Stress-Testing Two Approaches to LLM Reasoning Reliability: Test-Time Probability Aggregation and Logic-Representation Editing ​
Author: Minhan Cho, Jimin Kweon
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.08514v1 Announce Type: cross Abstract: We independently reproduce two recent methods for making large language model (LLM) reasoning more reliable, and stress-test them across domains and models (RPC across four new task domains with Qwen3-8B, LCF across four 7-8B models). The first, RPC,...
99. What Iterated Self-Feeding Probes of Language Models Measure, and a test that separates the construction from the model ​
Author: Nicol'as Vera Z'u~niga
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.10986v1 Announce Type: cross Abstract: A growing class of methods probes a language model by feeding it its own output: self-consistency, iterated refinement, agentic loops. We ask what such a probe measures, in a construction chosen to make the question sharp: a ring of token cells resam...
100. Detecting a Route Flip Is Easier Than Knowing Whether to Fix It: Causal Route-Mediated Damage in Quantized Mixture-of-Experts ​
Author: Parvel Gu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.11212v1 Announce Type: cross Abstract: Top-k Mixture-of-Experts (MoE) routing is discontinuous, so a deployment-motivated numerical disturbance -- simulated 4-bit KV-cache quantization read by a protected BF16 gate -- pushes tokens across decision boundaries and flips which experts fire. ...
101. Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop ​
Author: Igor Itkin
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cond-mat.stat-mech, cs.CL, cs.LG, cs.MA, physics.soc-ph
arXiv:2608.11215v1 Announce Type: cross Abstract: Simulating societies of many large language model (LLM) agents is expensive, yet the questions asked of such simulations are usually macroscopic: phase behaviour, stylised facts, and scaling with the number of agents $N$, not the cognition of any sin...
102. MaSRead: Content-Addressed Reading of Replicated Latent Stores ​
Author: Carlos Baquero, Lu'is Brito, Jo~ao Resende
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA
arXiv:2608.11218v1 Announce Type: cross Abstract: Independent agents that reason in latent space can share computed state as key-value cache fragments rather than text. Merged by a conflict-free replicated data type, these fragments form a store that converges under any delivery order or duplication...
103. From Monolithic to Modular: Segment-level Automatic Prompt Optimization ​
Author: Nikita Kulin, Viktor Zhuravlev, Artur Khairullin, Sergey Muravyov, Ilya Makarov, Daniil Sukhorukov, Ekaterina Averkova
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.11219v1 Announce Type: cross Abstract: Automatic Prompt Optimization (APO) often rewrites prompts monolithically, which can improve one behavior while degrading others. We present SAPO, a segment-level APO method that decomposes prompts into role, context, tasks, and output format, then a...
104. Transit Destination Inference from Tap-In-Only Bus Smart-Card Data: A Hierarchical Bayesian Approach ​
Author: Gefei Zhao, Jiahe Ling, Yuelong Su
Published: 8/13/2026, 4:00:00 AM
Categories: stat.AP, cs.LG, cs.SY, eess.SY, stat.ML
arXiv:2608.11223v1 Announce Type: cross Abstract: Entry-only automatic fare collection systems record boardings but not alightings, preventing direct construction of origin-destination (OD) matrices. This study develops a Hierarchical Bayesian Latent-Destination (HBLD) model that combines station-ho...
105. Forecasting Side Effects of Activation Steering ​
Author: Chong Yong Ong, Alson Wei Jie Sim, Peixin Zhang, Jun Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.11227v1 Announce Type: cross Abstract: Activation steering modifies a language model by adding a learned direction to its hidden activations, enabling targeted behavioral changes without retraining. While effective, steering often produces unintended side effects on other behaviors, makin...
106. Retrofitting Recurrent Depth into a Pretrained Language Model: Installation, Extrapolation, Transfer, and Retention at Two Parameter Budgets ​
Author: Mark Shapiro
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.11233v1 Announce Type: cross Abstract: A dense, pretrained language model can be retrofitted with recurrent depth and learn an iterative latent transition that persists after outcome-only annealing. Qwen2.5-0.5B-Instruct is split into a Prelude, a weight-tied Recurrent Block, and a Coda, ...
107. The Off-Support Barrier: Why Semantic Safety Constraints Are Not Learning-Problem Invariants, and What Follows for Prior Design, Containment, and Verification ​
Author: Yoshinori Watanabe
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.11243v1 Announce Type: cross Abstract: We argue that a single structural fact organizes a wide range of phenomena in contemporary AI safety: a semantic safety constraint (e.g., the agent does not escape its sandbox) is an off-support object. Formally, if q is the data distribution and (p...
108. Towards Sustainable Learning in Online Education: A Reinforcement Learning Approach ​
Author: Chaofan Zhai, Yicheng Song, Ravi Bapna, Junyao Ye
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG
arXiv:2608.11245v1 Announce Type: cross Abstract: Online education offers unprecedented scalability and accessibility to global learners from diverse backgrounds, but it often suffers from low engagement and poor long term learning effectiveness. To address these challenges, we introduce AI Tutor, a...
109. Towards the Harness of Embodied Agents ​
Author: Qi Wang, Tianyi Wang, Chengyang Li, Shikun Ban, Yurun Chen, Yizhong Ge, Jason Qin, Chengtai Li, Wentao Zhu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.RO
arXiv:2608.11246v1 Announce Type: cross Abstract: The success of coding agents has established the harness as a paradigm: what an agent achieves depends not on the model alone, but on the infrastructure around it. We ask whether the same paradigm extends to embodied agents in the physical world. We ...
110. Diffuse to Compress: Leveraging Diffusion LMs for Lossless Compression ​
Author: Angelo Nardone, Paolo Ferragina
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.IT, cs.LG, math.IT
arXiv:2608.11249v1 Announce Type: cross Abstract: We study the problem of lossless text compression, motivated by the rapid growth in the collection and storage of digital textual data - including plain text, source code, and structured formats such as XML - and by recent advances in neural language...
111. Variable Selection in the Context of AI Fairness ​
Author: Ivan Luciano Danesi, Chiara Frigerio, Fabio Maccaferri, Giorgio Alessandro Motta, Pietro Zecca
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG
arXiv:2608.11251v1 Announce Type: cross Abstract: Fairness in AI systems has become more important with recent regulatory demands, such as the EU AI Act. Traditional approaches often do not take into account philosophical ethics and social awareness. Variable selection processes, in particular, can ...
112. Symbolic Machine Learning for Vapor-Liquid Equilibrium Prediction in Cx-N2 Binary Mixtures ​
Author: Bongseok Kim, Suman Chakraborty, Gary Huang, Mehek Mathur, Guang Lin, Li Qiao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, physics.comp-ph
arXiv:2608.11255v1 Announce Type: cross Abstract: Accurate prediction of vapor--liquid equilibrium (VLE) for hydrocarbon-nitrogen mixtures remains challenging for cubic equations of state, particularly across broad ranges of composition and hydrocarbon chain length. While deep learning models can pr...
113. Temperature-Driven Sequential Modeling for the Prediction of Annual Power Conversion Efficiency Profiles of Organic Photovoltaic Materials: Douala Case Study ​
Author: Steve Cabrel Teguia Kouam, Rockefeller Rockefeller, Raoult Dabou Teukam, Jean-Pierre Tchapet Njafa, Patrick Sorrel Mvoto Kongo, Jean-Pierre Nguenang, Serge Guy Nana Engo
Published: 8/13/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG, physics.chem-ph
arXiv:2608.11261v1 Announce Type: cross Abstract: Organic photovoltaic (OPV) materials are promising candidates for distributed solar energy in tropical regions, yet existing virtual screening tools report static power conversion efficiency (PCE) values at standard testing conditions (STC) that fail...
114. CosMAP: Contrastive Manifold Approximation and Projection for Dimensionality Reduction of Omics and Genealogical Data ​
Author: Fenosoa Randrianjatovo, Maya Saleh, Simon Girard, Amadou Barry
Published: 8/13/2026, 4:00:00 AM
Categories: q-bio.GN, cs.LG, stat.CO, stat.ME
arXiv:2608.11269v1 Announce Type: cross Abstract: Omics datasets, particularly single-cell RNA sequencing data, are high-dimensional, sparse, noisy, and dominated by zero values, making faithful low-dimensional representation challenging. Existing dimensionality-reduction methods may distort local n...
115. Hardware-Aware Deployment of Joint SAR Compression and Despeckling on FPGA ​
Author: C'edric L'eonard, Francescopaolo Sica, Martin Schulz
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.LG
arXiv:2608.11271v1 Announce Type: cross Abstract: Next-generation Synthetic Aperture Radar (SAR) missions will generate data far faster than they can downlink, making onboard data reduction essential for near-real-time Earth observation. Learned Image Compression (LIC) offers better rate-distortion ...
116. Uncertainty-Aware and Explainable Ensemble Deep Learning Framework for Multi-Class Skin Lesion Classification ​
Author: Rofiqul Islam, Lilatul Ferdouse
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG
arXiv:2608.11280v1 Announce Type: cross Abstract: Skin cancer diagnosis from dermoscopic images remains challenging due to high intra-class variability, inter-class similarity, class imbalance, and the limited interpretability of deep learning models. This paper proposes an uncertainty-aware and exp...
117. Physics-Informed Implicit Neural Representations for Improved Myocardial Perfusion MRI Quantification ​
Author: Christos Tsepas, Chang Yan, Maximilian Fuetterer, Sebastian Kozerke, Cian M Scannell
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.LG
arXiv:2608.11282v1 Announce Type: cross Abstract: Quantifying myocardial perfusion from cardiac magnetic resonance (CMR) can be achieved by fitting tracer-kinetic models to the dynamic contrast-enhanced MR data. However, fitting the observed data with multi-compartment exchange models, which describ...
118. SegPAR: Class-Centric Decision-Based Sparse Attack for Semantic Segmentation ​
Author: Dongsu Song, DaeYun GO, Boseung Seo, Jay Hoon Jung
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CR, cs.LG
arXiv:2608.11285v1 Announce Type: cross Abstract: Despite the practical relevance of sparse decision-based black-box threats, they have received limited attention in semantic segmentation. To bridge this gap, we adapt the most representative decision-based black-box sparse attacks from the classific...
119. Benchmarking Cyberattack Detection in Electric Vehicle Charging Infrastructure with Benign User Updates ​
Author: Hannan Chen, Roshni Anna Jacob, Jie Zhang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.SY, eess.SY
arXiv:2608.11286v1 Announce Type: cross Abstract: Cyberattack detection in electric vehicle charging infrastructure is complicated by legitimate post-activation revisions to requested energy and departure time. Charging manipulation attacks can exploit the same interface and variables; therefore, de...
120. CLEAR: Class-wise Expert Aggregation with Structured Sampling for Long-Tailed Classification ​
Author: Gawon Lim
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.11287v1 Announce Type: cross Abstract: Long-tailed classification poses a reliability challenge because models trained on imbalanced data are unevenly reliable across frequent and underrepresented classes. While existing methods address imbalance through re-balancing, adjustment, represen...
121. Uncertainty-Aware Compositional Localization and Placement Assessment of Catheters and Tubes in Chest X-Rays ​
Author: Harshil Lodhiya
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2608.11288v1 Announce Type: cross Abstract: Assessing catheter and tube placement on chest X-rays is safety-critical yet tedious and error-prone. Current deep learning methods either classify placement globally -- losing track of which device is where -- or segment all devices into a single ma...
122. Dueling Deep Q-Learning for Intrusion Detection ​
Author: Logan Luna (Georgia Institute of Technology), Matthew P. Berkowitz (Embry-Riddle Aeronautical University), Laxima Niure Kandel (Embry-Riddle Aeronautical University), Sirio Jansen-S'anchez (Embry-Riddle Aeronautical University)
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.11291v1 Announce Type: cross Abstract: Intrusion detection systems (IDS) and automated systems for detecting and reporting cyber threats, are commonly handled via supervised machine learning methods. Though effective, these models struggle to effectively adapt to new attack types. This st...
123. Clinical Feasibility of Low-Magnification Fluorescence Imaging for Breast Cancer Margin Detection Using Texture Analysis and Deep Learning ​
Author: Pouya Afshin, Tianling Niu, Tongtong Lu, David Helminiak, Julie Jorns, Mollie Patton, Tina Yen, Donghye Ye, Bing Yu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.11317v1 Announce Type: cross Abstract: High-resolution images of unprocessed surgical breast tissue can be obtained using microscopy with ultraviolet surface excitation (MUSE). This technique is considered a promising method for checking surgical margins during breast cancer surgery. In t...
124. Spectral graph clustering with inhomogeneous latent geometry ​
Author: Konstantin Avrachenkov, Lucas S. Sibemberg, Alexander Van Werde
Published: 8/13/2026, 4:00:00 AM
Categories: cs.SI, cs.LG, math.PR, stat.ML
arXiv:2608.11321v1 Announce Type: cross Abstract: We study spectral clustering in the presence of a confounding latent geometry. The leading eigenvectors may then be dominated by the latent geometry rather than by the communities. Nevertheless, we show in a block latent-space model that communities ...
125. Deployment Decision Reliability: A Generalizability-Theory Framework for Sizing Long-Horizon Agent Evaluations ​
Author: Vasundra Srinivasan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.11323v1 Announce Type: cross Abstract: Enterprise practitioners read agent leaderboards as if they ranked agent capability. We show, across three open agent-trace benchmarks (TheAgentCompany, $\tau^2$-bench, and AppWorld), that the agent main effect accounts for less than 3% of total vari...
126. Better, Faster, Stronger: Programmatic Skill Learning Best Reduces Agent Cost ​
Author: Zixi Huang, Xiheng Wang, Andrew Wang, William Jurayj, Bernal Jim'enez Guti'errez, Daniel Khashabi, Nicholas Andrews
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.11338v1 Announce Type: cross Abstract: Recently, the practice of augmenting LLM agent capability with skills has gained prevalence. We explore the cost effective adaptation of agents to novel domains by means of learning skills. Existing works focus on performance gain over cost effective...
127. Can Frontier LLMs Match Natively Multimodal Embeddings? A Comparison on Hard-Negative Text-to-Image Retrieval ​
Author: Archan Dutta, Vyanktesh Kanungo
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.IR, cs.LG
arXiv:2608.11343v1 Announce Type: cross Abstract: Multimodal retrieval and classification across different types of media, spanning text, images,video and audio, has traditionally relied on dual-encoder models that align visual and textual representations through contrastive learning. The March 2026...
128. ODE-Based Transformer Decoders for Iterative Sign Language Translation ​
Author: Tu\u{g}\c{c}e K{\i}z{\i}ltepe, Hacer Yalim Keles
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.11352v1 Announce Type: cross Abstract: Sign language translation has achieved strong results with Transformer architectures, yet recent improvements largely rely on scaling model capacity at the cost of increased computation. We propose a parameter-efficient alternative that improves expr...
129. Adaptation of Generalist Robot Policies with Minimal Data ​
Author: Shreyas Kowshik, Sreyas Venkataraman, Leo Wang, Niharika Pant, Max Simchowitz, Aviral Kumar
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.11363v1 Announce Type: cross Abstract: A central goal in robot learning is to move beyond task-specific human data collection toward robots that improve through autonomous interaction. Yet fully autonomous learning remains difficult with current policies: sparse rewards and weak zero-shot...
130. From Numbers to Judgment: Specialist LLM Agents and Reinforcement Learning for European Listed Real Estate ​
Author: Pardis Taghavi, Santosh Bhavani
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.11381v1 Announce Type: cross Abstract: We study whether the localized numerical operations and integrative judgments of financial analysis benefit from the same form of LLM specialization. Larix maps a 16-lens European listed-real-estate analysis framework to eight lens-aligned specialist...
131. Generative Learning for Quantum Measurement Design ​
Author: Jun Dai, Olivier Nahman-L'{e}vesque, Guillaume Rabusseau, Hong-Ye Hu, Cunlu Zhou
Published: 8/13/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2608.11396v1 Announce Type: cross Abstract: Extracting quantum information from a quantum state is a fundamental task of quantum computation, often requiring the estimation of many non-commuting observables under a finite measurement budget. For both near-term and early fault-tolerant settings...
132. When Self-Consistency Backfires: Majority Vote Hurts the Majority of Hard Science Problems for Small LLMs ​
Author: Utkarsh Bahuguna
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2608.11403v1 Announce Type: cross Abstract: Self-consistency (SC) via majority vote is a widely used way to spend inference-time compute: sample N chains of thought, return the plurality answer. On the full GPQA Diamond benchmark (198 graduate-level science questions), majority voting reduces ...
133. Large-scale AI-Ready Data for Anti-Cancer Drug Response Modeling ​
Author: Vincent Lavelle, Yitan Zhu, Kaitlyn Marlor, Thomas Brettin, Rick Stevens
Published: 8/13/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG
arXiv:2608.11444v1 Announce Type: cross Abstract: Drug response prediction (DRP) models are an active area of research in pharmacogenomics, with growing potential to accelerate the identification of effective anticancer drugs. However, their predictive performance is often constrained by limited dat...
134. Gaussian Meta-Space Augmentation for Stacking Ensembles in Multimodal IPMN Risk Stratification ​
Author: Max A. Nelson, Eminenur Sen Tasci, Zhixiang Wang, Zongwei Zhou, Halil Ertugrul Aktas, Andrea M. Bejar, Elif Keles, Ziliang Hong, S{\i}tk{\i} Safa Taflan, Muhammed Enes Tasci, Frank H. Miller, Michael B. Wallace, Rajesh N. Keswani, Gorkem Durak, Ulas Bagci
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.11472v1 Announce Type: cross Abstract: Pancreatic cancer is among the most lethal malignancies; risk stratification of intraductal papillary mucinous neoplasms (IPMNs) offers a crucial opportunity for early intervention but typically requires invasive tissue biopsy. Dominant vision-based ...
135. Probing and steering biology across Boltz-1s trunk-diffusion boundary ​
Author: Piotr Jedryszek, Tongmeng Xie, Adam Winnifrith, Alexander Hasson, Weronika 'Slesak, George Wicks, Toby Winnifrith, Oliver M. Crook
Published: 8/13/2026, 4:00:00 AM
Categories: q-bio.QM, cs.LG
arXiv:2608.11475v1 Announce Type: cross Abstract: AlphaFold3-class structure predictors pair a representational trunk, which processes sequence and context, with a diffusion module, which generates atomic coordinates. How biological information changes as it crosses this architectural boundary remai...
136. Forward Trajectory Steering for Hamilton-Jacobi Reachability Analysis ​
Author: Sungje Park, Stephen Tu
Published: 8/13/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2608.11480v1 Announce Type: cross Abstract: Hamilton-Jacobi (HJ) reachability provides a mathematically rigorous framework for safe control of dynamical systems, but its practical application is bottlenecked by the computational complexity of solving Hamilton-Jacobi-Isaacs variational inequali...
137. A Modular Agentic Framework for Synthetically Constrained Multi-Objective Hit-to-Lead Optimization ​
Author: Kelvin P. Idanwekhai, Enes Kelestemur, Benjamin Strickland, Matthew Hart, Steini Davidsson, Angelos Angelopoulos, Ron Alterovitz, Marcello DeLuca, Alexander Tropsha
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-bio.QM
arXiv:2608.11483v1 Announce Type: cross Abstract: Hit-to-lead optimization requires iterative design of hit analogs across competing potency, selectivity, physicochemical, pharmacokinetic, safety, and synthetic constraints. We present SABLE (Synthetically-accessible Agentic Bayesian Ligand Explorati...
138. Cross-Corpus Evaluation of Generalizable Vulnerability Detection in IoT Firmware ​
Author: Sadib Hassan Rumman, Md. Shariful Islam, Md. Rayhanur Rahman
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2608.11492v1 Announce Type: cross Abstract: IoT firmware vulnerability detection remains challenging due to heterogeneous firmware ecosystems, resource-constrained platforms, and limitations in existing benchmarks. Many datasets are synthetic or general-purpose and lack human-verified, contami...
139. From Prompting to Behavioral Alignment: Personalized LLM Judges for Recommendation Evaluation ​
Author: Alireza S. Ziabari, Kat Ellis, Colleen Chan, Ding Tong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.11493v1 Announce Type: cross Abstract: Traditional offline recommendation evaluation relies heavily on complex, manually maintained feature pipelines that are difficult to scale. While Large Language Models (LLMs) offer a promising alternative by predicting user engagement directly from r...
140. Language-Structured Relational Q-Learning for Threat-Aware Control in Safety-Critical Driving ​
Author: Aditya Humnabadkar, Huaizhong Zhang, Ardhendu Behera
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.ET, cs.LG
arXiv:2608.11498v1 Announce Type: cross Abstract: Natural-language-based scenario generation offers an intuitive means of describing rare and complex driving interactions, yet it is still uncertain whether training with language-structured data leads to truly adaptive control policies. We propose La...
141. Generative Semantic Segmentation via an Observable Semantic-Image Interface and Hierarchical Generator Evidence Alignment ​
Author: Weize Cai, Yongqi Dong, Zhida Shao, Zixin Fu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV
arXiv:2608.11537v1 Announce Type: cross Abstract: Generative semantic segmentation exposes structured predictions as images, but direct color decoding is susceptible to color drift and boundary mixing, whereas latent-feature decoders that predict a separate output distribution may relegate the rende...
142. Fine-Tuning Generative Models for Extreme Events via CVaR-Penalized Wasserstein Gradient Flows ​
Author: Thejani Gamage, Hyemin Gu, Zhizhen Zhang, Ziyu Chen, Markos Katsoulakis, Luc Rey-Bellet
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2608.11544v1 Announce Type: cross Abstract: We propose CVaR-penalized Generative Particle Algorithm (CVaR-GPA), a robust, tail-agnostic algorithm for fine-tuning generative models to learn heavy-tailed distributions and capture extreme events, requiring no prior knowledge or estimation of the ...
143. Beyond Single-Turn Confidence: Trajectory-Adapted Uncertainty Quantification for LLM Agents ​
Author: Dylan Bouchard, Mohit Singh Chauhan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.11552v1 Announce Type: cross Abstract: Uncertainty quantification (UQ) methods for language models are typically evaluated on single-turn outputs, where uncertainty is attached to one generated answer. For LLM agents, however, the unit of observation is an interactive trajectory, where th...
144. Unifying Physical Backpropagation ​
Author: Cyrill B"osch, Yigithan Gediz, Hakan T"ureci
Published: 8/13/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cs.ET, cs.LG, physics.optics
arXiv:2608.11585v1 Announce Type: cross Abstract: Physical computing systems exploit device dynamics for computation, but their gradient-based optimization is challenging: backpropagation through a digital twin suffers from model-reality gap. On-device gradient computation could resolve this issue, ...
145. Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning ​
Author: Xulin Fan, Jialu Li, Mohammad Nur Hossain Khan, Kexin Hu, Bashima Islam, Mark Hasegawa-Johnson, Nancy L. McElwain
Published: 8/13/2026, 4:00:00 AM
Categories: eess.AS, cs.CL, cs.LG
arXiv:2608.11587v1 Announce Type: cross Abstract: Recent advances in model design and self-supervised audio representations have improved speech and audio understanding, yet infant-centered naturalistic recordings remain challenging due to limited labeled data, low signal-to-noise ratio, and cross-f...
146. CookVoice: Unified Framework for Style Controllable Multi-Modal Human Voice Generation ​
Author: Haowei Lou, Hye-Young Paik, Dai Jia, Kai Li, Lina Yao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.SD, cs.LG
arXiv:2608.11590v2 Announce Type: cross Abstract: Human voice generation has made rapid progress in speech generation, singing voice generation, voice cloning, and voice editing. However, most existing systems are designed for specific tasks and often rely on task-dependent architectures, control si...
147. IoT-Enabled Autonomous Maritime Navigation in Smart Ports: A Curriculum-Guided Shared Policy Learning Framework ​
Author: Yuqing Lin, Rangya Zhang, Kum Fai Yuen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2608.11597v1 Announce Type: cross Abstract: As smart port infrastructures increasingly rely on autonomous maritime devices enabled by the Internet of Things (IoT), ensuring reliable onboard navigation intelligence has become a critical challenge for safe and scalable operations in congested wa...
148. MBA: Multimodal Benchmark and Agents for Real-World Business Ideation ​
Author: Hojun Choi, Jaeyo Shin, Suin Lee, Hyunjung Shim
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG
arXiv:2608.11616v2 Announce Type: cross Abstract: Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation. Yet existing approaches remain confined to a text-only paradigm, despite the inherently multimodal nature of real-world contexts. We thus int...
149. CAM-Guided Saliency Cutout and Image-Based Malware Classification ​
Author: Yasaman Ebrahimi, Martin Jurecek, Mark Stamp
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.11634v1 Announce Type: cross Abstract: Dropout regularization is commonly used to reduce overfitting by removing parts of a neural network during training. For Convolutional Neural Networks (CNN), cutouts serve a somewhat analogous purpose. Cutouts can be implemented as data augmentation:...
150. Robustness of AI-Art Detectors under Generator Shift ​
Author: Shivank Singh Thakur, Meien Li, Mark Stamp
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.11643v1 Announce Type: cross Abstract: Text-to-image generative models have advanced rapidly, with modern Diffusion Transformer architectures producing images that are increasingly difficult to distinguish from human-created artwork. This development has raised significant concerns regard...
151. A Quantum/Classical Example Oracle Separation for Making Things Up ​
Author: Kenny Chen
Published: 8/13/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, stat.ML
arXiv:2608.11648v1 Announce Type: cross Abstract: We study the power of quantum examples, as compared to classical examples, in the PAC learning framework. Here, we have two learning algorithms, both with access to quantum computation, but one gets quantum examples, whereas the other gets classical ...
152. Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing ​
Author: Tianci Liu, Zihan Dong, Tianchun Li, Yi-Chung Chen, Qiming Cao, Xingchen Wang, Shiyang Wang, Zichen Miao, Linjun Zhang, Haoyu Wang, Jing Gao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.11660v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable performance across natural language tasks, yet they are trained on static corpora and their knowledge quickly becomes outdated in a fast-changing world. This motivates knowledge editing (KE), which upda...
153. APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference ​
Author: Alish Kanani, Layan Badawi, Umit Y. Ogras
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AR, cs.AI, cs.LG
arXiv:2608.11688v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models are attractive for edge deployment because they provide high model capacity while activating only a small subset of parameters per token, improving compute efficiency. However, MoE inference at the edge is fundamentall...
154. Locating and Controlling Implicit Personalization in Large Language Models ​
Author: Yueru Yan, Siqi Wu, Thai Le
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.11735v1 Announce Type: cross Abstract: Large language models (LLMs) often shift their outputs in response to implicit demographic cues even when users never state a demographic identity. Previous work has documented this behavior, but the connection between these behavioral changes and th...
155. Automated binary classification of hazelnut X-ray images: A deep-learning benchmark for quality assessment ​
Author: Giancarlo Sportelli, Nicola Belcari, Roberta Pace, Umberto Bernardo, Sharmin Sultana, Alessandra Toncelli, Matteo Giaccone
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, physics.app-ph
arXiv:2608.11759v1 Announce Type: cross Abstract: Non-destructive X-ray imaging can reveal internal hazelnut defects that are difficult to detect by external inspection alone; however, automated interpretation remains challenging because of subtle radiographic differences among classes, marked class...
156. Tight Nonasymptotic Local Convergence of Sinkhorn-Knopp ​
Author: Wenzhi Gao, Zhaonan Qu, Yinyu Ye, Madeleine Udell
Published: 8/13/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ML
arXiv:2608.11760v2 Announce Type: cross Abstract: We revisit the Sinkhorn-Knopp (SK) algorithm for the matrix scaling problem. Despite extensive literature on the global convergence of SK and its variants, its local linear convergence behavior remains less understood. We address this gap by providin...
157. A comparison of CNN architectures for Alzheimer's disease detection in single-view MRI scans ​
Author: Hiram Zuniga, Ulises Orozco-Rosas, Kenia Picos
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2608.11762v1 Announce Type: cross Abstract: Alzheimer's disease is a leading cause of death with no cure. Therefore, early detection is critical to slow progression and preserve quality of life. Diagnosis relies on medical history, cognitive tests, physical exams, and MRI brain scans, making d...
158. Achieving Near-Zero-Overhead Multi-Model Hierarchical Classification in Real-Time Detection Pipelines ​
Author: Vaishnav Raju
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.DC, cs.LG
arXiv:2608.11770v1 Announce Type: cross Abstract: Edge-deployed vision systems in target recognition, surveillance, autonomous vehicles, and drone domains require hierarchical inference pipelines where a detection model identifies objects of interest and downstream classifiers provide fine-grained a...
159. GRPO for Financial Advice Generation: Outperforming Commercial LLMs under CATE Evaluation ​
Author: Ofir Ben Shoham, Shrutendra Harsola, Vignesh Subrahmaniam, Shravan Mohan, Yakov Gazman, Oded Vainas
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.11787v1 Announce Type: cross Abstract: Generating actionable financial advice from business records demands that models integrate numerical reasoning, domain knowledge, and sound judgment, while avoiding recommendations that could harm the business. Direct supervision is difficult: histor...
160. Can Vision Models Read the Radar Display? On the Feasibility of Radar Imagery for Air Traffic Complexity Estimation ​
Author: Hyewook Kim, Byul Kang, Seokbin Yoon, Keumjin Lee
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.11810v1 Announce Type: cross Abstract: Air traffic controllers perceive traffic complexity through the radar display, suggesting that a computer vision model operating on the same imagery may provide a natural architecture for modeling controller-perceived complexity; however, whether rad...
161. Located but Not Releasable: Silent Gate Inversion and Bounded Linear Release ​
Author: Xining Xun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.11822v1 Announce Type: cross Abstract: A growing body of work reports that language models represent task-relevant latent structure that they fail to use. Whether such structure, once located, can be converted into behavior is a separate question that is rarely tested end to end. We submi...
162. User-Assisted Collaborative Distributed Inference for Efficient QoS-Aware Autoscaling ​
Author: Alfreds Lapkovskis, Ali Beikmohammadi, Sindri Magn'usson, Praveen Kumar Donta
Published: 8/13/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG, cs.NI, cs.PF
arXiv:2608.11840v1 Announce Type: cross Abstract: Growing demand for artificial intelligence (AI) inference services requires scalable infrastructure, yet centralized serving costs rise with demand. We propose a collaborative distributed inference system combining dedicated infrastructure with resou...
163. Policy-as-logic for robust reasoning over rules ​
Author: Rahul Nair, Bastian Lipka, Elizabeth Daly
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SC
arXiv:2608.11905v1 Announce Type: cross Abstract: In many practical applications of generative AI systems, from tax rules to airline baggage allowance, responses to natural language queries must respect written policies or rules. We present a hybrid symbolic approach that expresses policies in forma...
164. LODESTAR: Trustworthy Entropy Is Navigated, Not Merely Measured -- Reinforced Polarizer Keeps a Frozen LLM from Being Confidently Misled by the Wrong Evidence ​
Author: Po-Jen Ko, Che-Cheng Wu, Hung-Chun Hsu, Li-Yang Chang, Chuan-Ju Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.IR, cs.LG
arXiv:2608.11922v1 Announce Type: cross Abstract: Predictive-distribution entropy makes a strong selection rule in retrieval-augmented question answering: across five QA benchmarks, keeping the candidate answer that a frozen respondent LLM produces with the lowest answer-token entropy lifts mean ans...
165. Latent variable models for simultaneous EOV identification and removal in population-based SHM ​
Author: M. D. Champneys, M. R. Jones, A. J. Hughes, T. J. Rogers, E. J. Cross, K. Worden
Published: 8/13/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2608.11995v1 Announce Type: cross Abstract: The robust treatment of environmental and operational variability (EOV) is an open challenge in population-based structural health monitoring (PBSHM). The difficulty is compounded in the case that the EOV signals are unmeasured. A common approach in ...
166. A Remote Approach to Cashew Orchard Detection: Leveraging Active Learning with Satellite Imagery in Guinea-Bissau ​
Author: Miguel, Sofia, Maria, Patr'icia, Luke, Jo~ao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.11996v1 Announce Type: cross Abstract: Cashew production is a widespread economic activity in Guinea-Bissau, as well as other countries in West Africa. However, unregulated cashew production can be directly associated with increasing regionwide deforestation rates, biodiversity losses, an...
167. Beyond Local Power: Functional Connectivity Analysis for Subject-Independent Learning Style Recognition ​
Author: Wiga Maulana Baihaqi, Indriana Hidayah, Sri Kusrohmaniah, Noor Akhmad Setiawan
Published: 8/13/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG, eess.SP
arXiv:2608.12000v1 Announce Type: cross Abstract: Identifying individual learning styles optimizes pedagogical efficacy. While traditional questionnaires are structured, behavioral tracking methods require prolonged interaction log accumulation. To overcome these temporal constraints, this paper pro...
168. Adaptive Bregman Proximal Stochastic Gradient with a Stabilized Barzilai--Borwein Step Size ​
Author: Chenhan Jin, Shengze Xu, Binghui Xie, Kaiwen Zhou, Fan Jia, James Cheng, Tieyong Zeng
Published: 8/13/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2608.12009v1 Announce Type: cross Abstract: Bregman proximal stochastic gradient (BPSG) methods bring variance-reduced composite optimization to objectives whose geometry is poorly captured by Euclidean smoothness. Their performance, however, remains sensitive to the step size: raw stochastic ...
169. Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence ​
Author: Mengru Wang, Junfeng Fang, Shuofei Qiao, Zhenqian Xu, Haoming Xu, Haoxiong Wang, Shumin Deng, Linyi Yang, Zhixiang Cui, Xin Xu, Yunzhi Yao, Buqiang Xu, Fei Shen, Haozhe Luo, Yunxiang Wei, Ningyu Zhang, Julian McAuley, Tat Seng Chua, Huajun Chen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.HC, cs.LG, cs.MA
arXiv:2608.12036v1 Announce Type: cross Abstract: AI models have achieved remarkable success across diverse domains, yet the mechanisms underlying their capabilities and the risks they may pose remain poorly understood. As AI development becomes faster and increasingly automated, mechanistic explora...
170. Direct Acceleration of Stochastic Root-Finding Without Variance Reduction and Regularization ​
Author: TaeHo Yoon, Nicolas Loizou
Published: 8/13/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2608.12043v1 Announce Type: cross Abstract: Acceleration for deterministic root-finding problems has been extensively studied in recent years; specifically, the anchor-based, or Halpern-type methods achieve optimal convergence rates with respect to the operator norm. However, acceleration via ...
171. Draw This First ​
Author: Dazhi Zhong, Rowan Bradbury, Grant Davis
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.12064v1 Announce Type: cross Abstract: We invert the typical formulation of sketch generation: instead of drawing strokes in order, we predict a 2D field that defines the order in which strokes are drawn. We use a pretrained latent flow-matching transformer to supply the image prior to pr...
172. Better Slots, Better Worlds: Representation Quality & Robustness in Object-Centric World Models ​
Author: Shukrullo Nazirjonov, Sai Prasanna, Anna Manasyan, Georg Martius
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.12078v1 Announce Type: cross Abstract: Learning world models from offline trajectories enables agents to accomplish different tasks through planning. Object-centric (OC) representations, which decompose a scene into a set of slots that bind to its objects, have been proposed as an inducti...
173. Look What the Probes Dragged In! Real-World Chest X-ray Shortcuts in MedCLIP ​
Author: Nikolette Pedersen, Regitze Sydendal, Veronika Cheplygina, Th'eo Sourget
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.12086v1 Announce Type: cross Abstract: Vision-language models, such as contrastive language-image pre-training (CLIP)-based approaches, have reached state-of-the-art (SOTA) results in medical artificial intelligence. However, recent work reveals that CLIP-based models remain vulnerable to...
174. The Advective Fisher-Rao Geometry of Deterministic Measure Transport ​
Author: Benjamin Gess, Johannes M"uller
Published: 8/13/2026, 4:00:00 AM
Categories: math.OC, cs.LG, math.DG, math.PR
arXiv:2608.12111v1 Announce Type: cross Abstract: A novel advective Fisher-Rao metric is introduced for optimization tasks on paths of probability measures governed by the continuity equation. This metric is shown to lead to optimal descent directions. It is then shown that this metric arises natura...
175. A corpus-specific clinical RAG system matches or outperforms newer frontier LLMs on HealthBench ​
Author: Praveen Reddy, Charuta Mandke, Suvrankar Datta, Sarah Khan, Siddharth Reddy Anthireddy, Shitij Arora, Vishal Singh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC, cs.IR, cs.LG
arXiv:2608.12138v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) have recently been reported to match or exceed specialized clinical AI tools on medical benchmarks, but such comparisons draw on a narrow set of systems and on benchmarks developed largely in high-income s...
176. FQTree: Fine-grained Quantization and Hardware Generation of Boosted Decision Trees ​
Author: Zhiqiang Que, Chang Sun, Haiyang Wang, Dinesh Pamunuwa, Roshan Weerasekera, Qijia Tang, Bakhtiar Zadeh, Wayne Luk, Maria Spiropulu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2608.12140v1 Announce Type: cross Abstract: Boosted decision trees (BDTs) are widely used in latency-critical applications, but efficient hardware deployment remains challenging. Existing designs often rely on uniform or manually tuned fixed-point formats, which can introduce unnecessary hardw...
177. ADEPT: A Unified Framework for Deep Learning Test Adequacy ​
Author: Yidi Kao, Shawn Burnham, Tommi Rose Fahy, Ali Ghanbari
Published: 8/13/2026, 4:00:00 AM
Categories: cs.SE, cs.LG
arXiv:2608.12144v1 Announce Type: cross Abstract: Over the past decade, many test adequacy metrics have been proposed for deep learning that characterize test dataset adequacy from different perspectives, e.g., neuron activation behavior, latent feature coverage, decision-boundary exploration, etc. ...
178. Autonomous Telerehabilitation via Skeletal Motion Prediction and Joint-Level Performance Assessment ​
Author: Lara Pereira, Jo~ao Ruivo Paulo, Pedro Santos, Paulo Peixoto
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2608.12145v1 Announce Type: cross Abstract: Autonomous rehabilitation systems must not only recognize human motion but also provide structured feedback to support users without continuous therapist supervision. This paper presents a telerehabilitation pipeline that integrates skeleton-based ex...
179. How to Spend Your Oracle Budget: Practical Guidance for Protein Structure Prediction Models ​
Author: Aleksandra Kalisz, Jack Simons, Krisztina Sinkovics, Noam Ghenassia, Shikha Surana, Henry Moss, Paul Duckworth
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2608.12192v1 Announce Type: cross Abstract: Foundation models for protein structure prediction remain unreliable on certain targets. External oracles can flag and correct these failures, but biological oracles are expensive, making oracle budget a critical constraint. Existing guidance methods...
180. Learning-Based Behavior Planning for Automated Driving: Real-World Integration and Deployment ​
Author: Jean-Pierre Busch, Guido Linden, Jan Bergmann, Lutz Eckstein
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2608.12198v1 Announce Type: cross Abstract: Recent research in machine and deep learning has shown the potential of learningbased motion planning approaches to improve the driving behavior of automated vehicles, especially in complex environments. However, their complex nature and lack of tran...
181. Regime-Gated Residual Mixture-of-Experts for Cross-Sectional Volatility Forecasting ​
Author: Junyi Ye, Gargi Vijay Borde
Published: 8/13/2026, 4:00:00 AM
Categories: q-fin.ST, cs.LG
arXiv:2608.12251v1 Announce Type: cross Abstract: Financial volatility is regime dependent, yet incorporating regime information into neural networks can also destabilize training. This paper asks where such information should enter a neural cross-sectional volatility forecasting model. We study fiv...
182. One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL ​
Author: Simon Yu, Nicholas Tomlin, Marwa Abdulhai, Ximing Lu, Derek Chong, Abe Hou, Dilara Soylu, Sergey Levine, Christopher D. Manning, Weiyan Shi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that this approach systematically fails to generalize, and trace the failure to simulator collapse: becau...
183. Weak Correlations as the Underlying Principle for Linearization of Gradient-Based Learning Systems ​
Author: Ori Shem-Ur, Khen Cohen, Aviv Orly, Yaron Oz
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, hep-th, math.PR, stat.ML
arXiv:2401.04013v3 Announce Type: replace Abstract: Deep learning models, such as wide neural networks, can be conceptualized as nonlinear dynamical physical systems characterized by a multitude of interacting degrees of freedom. Such systems in the infinite limit, tend to exhibit simplified dynamic...
184. Deep Activity Model: A Generative Approach for Human Mobility Pattern Synthesis ​
Author: Xishun Liao, Qinhua Jiang, Brian Yueshuai He, Yifan Liu, Chenchen Kuai, Jiaqi Ma
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2405.17468v3 Announce Type: replace Abstract: Human mobility plays a crucial role in transportation, urban planning, and public health, but current approaches face important limitations. Existing deep learning models tend to overlook the semantic interdependencies among activities and househol...
185. A New First-Order Meta-Learning Algorithm with Convergence Guarantees ​
Author: El Mahdi Chayti, Martin Jaggi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2409.03682v2 Announce Type: replace Abstract: Learning new tasks by leveraging prior experience is a fundamental trait of intelligent systems. While Model-Agnostic Meta-Learning (MAML) is a leading approach, it suffers from significant computational and memory overhead due to the requirement o...
186. Program Semantic Inequivalence Game with Large Language Models ​
Author: Antonio Valerio Miceli-Barone, Vaishak Belle, Ali Payani
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.PL
arXiv:2505.03818v3 Announce Type: replace Abstract: Large Language Models (LLMs) can achieve strong performance on everyday coding tasks, but they can fail on complex tasks that require non-trivial reasoning about program semantics. Finding training examples to teach LLMs to solve these tasks can be...
187. Computational Algebra with Attention: Transformer Oracles for Border Basis Algorithms ​
Author: Hiroshi Kera, Nico Pelleriti, Yuki Ishihara, Max Zimmer, Sebastian Pokutta
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.SC
arXiv:2505.23696v2 Announce Type: replace Abstract: Solving systems of polynomial equations, particularly those with finitely many solutions, is a crucial challenge across many scientific fields. Traditional methods like Gr"obner and Border bases are fundamental but suffer from high computational c...
188. Improving Performance of Spike-based Deep Q-Learning using Ternary Neurons ​
Author: Aref Ghoreishee, Abhishek Mishra, John Walsh, Anup Das, Nagarajan Kandasamy
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.NE, cs.SY, eess.SY
arXiv:2506.03392v3 Announce Type: replace Abstract: We propose a new ternary spiking neuron model to improve the representation capacity of binary spiking neurons in deep Q-learning. Although a ternary neuron model has recently been introduced to overcome the limited representation capacity offered ...
189. Learning Multi-Timescale Interventions under Safety and Resource Constraints ​
Author: David Mguni, Wanrong Yang, Jing Dong, Jing Peng, Ziquan Liu, Muhammad Salman Haleem, Baoxiang Wang, Dominik Wojtczak
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2508.03875v3 Announce Type: replace Abstract: Many sequential decision problems offer qualitatively different ways of influencing the environment: some interventions act immediately, whereas others induce persistent effects that continue to shape future states long after the decision that init...
190. Local Cluster Cardinality Estimation for Adaptive Mean Shift ​
Author: 'Etienne Pepin
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2508.12450v2 Announce Type: replace Abstract: This article presents an adaptive mean shift algorithm in which every parameter used at a point is derived from that point's own distance distribution. The distance distribution from a point to all others is used to estimate the cardinality of the ...
191. Patch-based Memory Gate Model in Time Series Foundation Model ​
Author: Samuel Yoon, Jongwon Kim, Juyoung Ha, Young Myoung Ko
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.18751v4 Announce Type: replace Abstract: Recently reconstruction-based deep models have been widely used for time series anomaly detection, but as their capacity and generalization capability increase, these models tend to over-generalize, often reconstructing unseen anomalies accurately....
192. Variational Mixture of Graph Neural Experts for Alzheimer's Disease Recognition across Frequency Bands in EEG Brain Networks ​
Author: Jun-En Ding, Anna Zilverstand, Shihao Yang, Albert Chih-Chieh Yang, Feng Liu
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.11917v4 Announce Type: replace Abstract: Dementia disorders such as Alzheimer's disease (AD) and frontotemporal dementia (FTD) exhibit overlapping electrophysiological signatures in electroencephalography (EEG) that challenge accurate diagnosis. Existing EEG-based methods are limited by f...
193. Adaptive Online Learning with LSTM Networks for Energy Price Prediction ​
Author: Salih Salihoglu, Ibrahim Ahmed, Afshin Asadi
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.16898v2 Announce Type: replace Abstract: Accurate prediction of electricity prices is crucial for stakeholders in the energy market, particularly for grid operators, energy producers, and consumers. This study focuses on developing a predictive model leveraging Long Short-Term Memory (LST...
194. Reliable Inference in Edge-Cloud Model Cascades via Conformal Alignment ​
Author: Jiayi Huang, Sangwoo Park, Nicola Paoletti, Osvaldo Simeone
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, stat.ML
arXiv:2510.17543v3 Announce Type: replace Abstract: Edge intelligence enables low-latency inference via compact on-device models, but assuring reliability remains challenging. We study edge-cloud cascades that must preserve conditional coverage: whenever the edge returns a prediction set, it should ...
195. LoRAQuant: Mixed-Precision Quantization of LoRA to Ultra-Low Bits ​
Author: Amir Reza Mirzaei, Yuqiao Wen, Yanshuai Cao, Lili Mou
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.26690v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has become a popular technique for parameter-efficient fine-tuning of large language models (LLMs). In many real-world scenarios, multiple adapters are loaded simultaneously to enable LLM customization for personalized us...
196. Accelerating Time Series Foundation Models with Speculative Decoding ​
Author: Pranav Subbaraman, Fang Sun, Jinxi Yu, Yue Yao, Huacong Tang, Xiao Luo, Yizhou Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.18191v2 Announce Type: replace Abstract: Time series forecasting drives operational decisions under tight latency budgets, and autoregressive time series foundation models (TSFMs) increasingly deliver the most accurate forecasts. That accuracy is paid for at inference, since a horizon of ...
197. BrowseSafe: Understanding and Preventing Prompt Injection Within AI Browser Agents ​
Author: Kaiyuan Zhang, Mark Tenenholtz, Kyle Polley, Jerry Ma, Denis Yarats, Ninghui Li
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2511.20597v2 Announce Type: replace Abstract: The integration of artificial intelligence (AI) agents into web browsers introduces security challenges that go beyond traditional web application threat models. Prior work has identified prompt injection as a new attack vector for web agents, yet ...
198. A-3PO: Accelerating Asynchronous LLM Training with Staleness-aware Proximal Policy Approximation ​
Author: Xiaocan Li, Shiliang Wu, Zheng Shen
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC
arXiv:2512.06547v4 Announce Type: replace Abstract: Decoupled PPO has been a successful reinforcement learning (RL) algorithm to deal with the high data staleness under the asynchronous RL setting. Decoupled loss used in decoupled PPO improves coupled-loss style of algorithms' (e.g., standard PPO, G...
199. Grounding Large Language Models as Generalizable Policies in Network Control ​
Author: Duo Wu, Linjia Kang, Zhimin Wang, Fangxin Wang, Wei Zhang, Chongbo Sun, Xuefeng Tao, Wei Yang, Le Zhang, Wenwu Zhu, Peng Cui, Zhi Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2512.11839v2 Announce Type: replace Abstract: Designing generalizable control policies that operate reliably under changing conditions is essential for robust network services in modern digital infrastructure. Yet network control remains dominated by specialized policies built from handcrafted...
200. Federated Learning for the Design of Parametric Insurance Indices under Heterogeneous Renewable Production Losses ​
Author: Fallou Niakh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2601.12178v2 Announce Type: replace Abstract: We propose a federated learning framework for the calibration of parametric insurance indices under heterogeneous renewable energy production losses. Producers locally model their losses using Tweedie generalized linear models and private data, whi...
201. Probably Approximately Correct Maximum A Posteriori Inference ​
Author: Matthew Shorvon, Frederik Mallmann-Trenn, David S. Watson
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2601.16083v2 Announce Type: replace Abstract: Computing the conditional mode of a distribution, better known as the maximum a posteriori (MAP) assignment, is a fundamental task in probabilistic inference. However, MAP is generally intractable, and remains hard even under many common structural...
202. Superposition Without Interference? Towards Isolated Interventions via Almost Orthogonal Features in Language Models ​
Author: Moritz Miller, Florent Draye, Bernhard Sch"olkopf
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2602.04718v5 Announce Type: replace Abstract: A central premise in mechanistic interpretability is that meaningful concepts in language models are represented by linear features in activation space. For such features to support reliable interventions, manipulating one feature should not substa...
203. A Theoretical Framework for Modular Learning of Robust Generative Models ​
Author: Corinna Cortes, Mehryar Mohri, Yutao Zhong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2602.17554v3 Announce Type: replace Abstract: Training large-scale generative models is resource-intensive and relies heavily on heuristic dataset weighting. We address two fundamental questions: Can we train Large Language Models (LLMs) modularly, combining small, domain-specific experts to m...
204. Representation Finetuning for Continual Learning ​
Author: Haihua Luo, Xuming Ran, Tommi K"arkk"ainen, Huiyan Xue, Zhonghua Chen, Qi Xu, Fengyu Cong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.11201v3 Announce Type: replace Abstract: The world is inherently dynamic, and continual learning aims to enable models to adapt to ever-evolving data streams. While pre-trained models have shown powerful performance in continual learning, they still require finetuning to adapt effectively...
205. Stochastic Dimension Zeroth-Order Estimator: Stable and Memory-Efficient Training of PINNs ​
Author: Zhangyong Liang, Huanhuan Gao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.24002v4 Announce Type: replace Abstract: Physics-Informed Neural Networks (PINNs) for high-dimensional and high-order partial differential equations (PDEs) are primarily constrained by the $\mathcal{O}(d^k)$ spatial derivative complexity and the $\mathcal{O}(P)$ memory overhead of backpro...
206. Weaves, Wires, and Morphisms: Formalizing and Implementing the Algebra of Deep Learning ​
Author: Vincent Abbott, Gioele Zardini
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, math.CT
arXiv:2604.07242v3 Announce Type: replace Abstract: Despite deep learning models running well-defined mathematical functions, we lack a formal mathematical framework for describing model architectures. Ad-hoc notation, diagrams, and pseudocode poorly handle nonlinear broadcasting and the relationshi...
207. Spend Less, Fit Better: Budget-Efficient Scaling Law Fitting via Active Experiment Selection ​
Author: Sijie Li, Shanda Li, Haowei Lin, Weiwei Sun, Ameet Talwalkar, Yiming Yang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.22753v2 Announce Type: replace Abstract: Scaling laws are used to plan multi-million-dollar training runs, but fitting those laws can itself cost millions. In modern large-scale workflows, assembling a sufficiently informative set of pilot experiments is already a major budget-allocation ...
208. Optimized Deferral for Imbalanced Settings ​
Author: Corinna Cortes, Anqi Mao, Mehryar Mohri, Yutao Zhong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2604.27723v2 Announce Type: replace Abstract: Learning algorithms can be significantly improved by routing complex or uncertain inputs to specialized experts, balancing accuracy with computational cost. This approach, known as learning to defer, is essential in domains like natural language ge...
209. Mind the Gap: Structure-Aware Consistency in Preference Learning ​
Author: Mehryar Mohri, Yutao Zhong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2604.27733v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human intent, whether through explicit reward modeling or direct methods such as DPO, fundamentally relies on minimizing a surrogate loss as a proxy for the true pairwise ranking objective. We prove that t...
210. Linear-Core Surrogates: Smooth Loss Functions with Linear Rates for Classification and Structured Prediction ​
Author: Mehryar Mohri, Yutao Zhong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2604.27742v2 Announce Type: replace Abstract: A fundamental dichotomy in the theory of classification sets smoothness against statistical efficiency: smooth surrogate losses such as the logistic loss enable fast $O(1/T)$ optimization but yield slow square-root $H$-consistency bounds, while pie...
211. Analytic Bridge Diffusions for Controlled Path Generation ​
Author: Michael Chertkov
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, cs.AI, cs.SY, eess.SY, math.OC
arXiv:2605.02961v2 Announce Type: replace Abstract: Most modern bridge-diffusion methods achieve finite-time transport by specifying an interpolation, Schrodinger-bridge, or stochastic-control objective and then learning the associated score or drift field with a neural network. In contrast, we iden...
212. Pretraining large language models with MXFP4 on Native FP4 Hardware ​
Author: Musa Cim, Sarthak Arora, Poovaiah Palangappa, Miro Hodak, Ravi Dwivedula, Meena Arunachalam, Mahmut Taylan Kandemir
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.09825v4 Announce Type: replace Abstract: Why does full-pipeline FP4 training of large language models often diverge, even when forward activations and activation gradients remain stable? We address this question through a controlled study of MXFP4 quantization in transformer training, pro...
213. Modeling Spectral Energy Shifts in Spatio-Temporal Graph Anomaly Detection ​
Author: Yilin Liu, Hongchao Zhang, Taylor T. Johnson, Ahmad F. Taha, Meiyi Ma
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.00304v2 Announce Type: replace Abstract: Graph anomaly detection methods aim to distinguish anomalous nodes. While prior methods characterize anomalies through increased variation in the spectral energy distributions, they overlook those that result in decreased variation, i.e., camouflag...
214. Planar Symmetric Pattern Generation ​
Author: Ning Lin, Luxi Chen, Huaguan Chen, Jiacheng Cen, Chongxuan Li, Wenbing Huang, Hao Sun
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.02073v2 Announce Type: replace Abstract: Generating objects with specific symmetries is essential in various real-world scenarios. However, adapting existing 2D continuous representations to enforce planar group symmetry remains a challenge, as the transformation of non-reflective group e...
215. Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement Learning ​
Author: Xuekang Wang, Zhuoyuan Hao, Shuo Hou, Hao Peng, Juanzi Li, Xiaozhi Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2606.04923v2 Announce Type: replace Abstract: Rubric-based reinforcement learning (RL) uses an LLM-as-a-Judge (LaaJ) to score model outputs according to rubrics as rewards. However, policy models may exploit latent biases in the judge, leading to reward hacking and ineffective or unsafe traini...
216. Bootstrap Theory of Representational Emergence: Explanatory Insufficiency as a Driver of Representation Learning and World Models ​
Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.07303v3 Announce Type: replace Abstract: Representation learning is central to modern machine learning, yet most research focuses on optimizing representations after a representational framework has been selected. Less attention is given to when a new representational level becomes necess...
217. Detecting Explanatory Insufficiency in Learned Representations: A Framework for Representational Vigilance ​
Author: Jacques Raynal, Pierre Slangen, Elsa Raynal, Jacques Margerit
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.13172v3 Announce Type: replace Abstract: Learned representations are central to modern machine learning and are typically evaluated through predictive performance, robustness, uncertainty estimation, and generalization. However, a learned representation may remain operationally successful...
218. ATMA: Long-Context Language Modeling via Polar Attention and Gated-Delta Compression Memory ​
Author: Habibullah Akbar
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.25156v4 Announce Type: replace Abstract: Length extrapolation in language models involves competing objectives: retrieval fidelity, long-document likelihood, short-context quality, and inference cost. We present ATMA, a 378M-parameter hybrid recipe that combines Polar Attention with gated...
219. Prompt-Driven Exploration ​
Author: Sunshine Jiang, John Marangola, David Zhang, Raghuram Kowdeed, Ruiyang Luo, Nitish Dashora, Richard Li, Pulkit Agrawal, Zhang-Wei Hong
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.08837v2 Announce Type: replace Abstract: Exploration is essential to RL since a policy cannot improve by repeatedly sampling the behaviors it already prefers. Standard methods inject stochasticity in the action space, but such jitter only yields rollouts close to the original. Escaping a ...
220. Quantum Port-Hamiltonian Neural Networks: Learning Conservative and Dissipative Dynamics via Measurement-Induced Nonlinearity ​
Author: Dibakar Sigdel
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.12269v3 Announce Type: replace Abstract: We introduce Quantum Port-Hamiltonian Neural Networks (Q-pHNNs), parameterised quantum circuits that learn classical dynamics in a structure-preserving manner. The framework rests on the Isomorphic Hamiltonian Mapping (IHM): the skew-symmetric inte...
221. Gauge-Fixing the Forward-Forward Objective: A Whitened Goodness Derived from a Likelihood-Ratio Account ​
Author: Paolo Giannitrapani
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, eess.IV, stat.ML
arXiv:2607.12501v3 Announce Type: replace Abstract: The Forward-Forward algorithm trains each layer locally, so that a scalar goodness - the sum of squared activations - is high on real inputs and low on contrastive ones. Under an explicit generative model this goodness is the sufficient statistic o...
222. Reducing Per-Sample Interference in Stochastic Optimization ​
Author: Apostolos Avranas
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.16261v2 Announce Type: replace Abstract: Modern optimizers combine gradients from the current mini-batch with historical optimization state, such as momentum or adaptive moments. While effective, this standard practice can produce parameter updates that actively increase the loss of indiv...
223. Agentic Bayesian Optimization through Surrogate-Augmented Autoresearch ​
Author: Paul Brunzema, Louis Tiao, Nhat Le, Kevin De Angeli, Yao Xuan, Djordje Gligorijevic
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.00316v2 Announce Type: replace Abstract: Bayesian optimization (BO) has become the standard tool for sample-efficient optimization and owes its efficiency to uncertainty-aware search driven by generic statistical priors. Richer domain priors can improve BO in principle, but encoding them ...
224. Empowering Credit Risk Detection in Weixin Pay with Billion-Scale Deep Graph Learning ​
Author: Xin Liu, Xiyuan Chen, Chenglong Wu, Xuan Zong, Jun Zhou, Dawei Cheng
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.02168v2 Announce Type: replace Abstract: Credit risk detection, particularly mitigating individual fraud, is crucial for maintaining the stability of digital financial ecosystems. Accurately identifying credit fraud among billions of users is critical for minimizing financial losses and s...
225. Unscented KalmanNet: Structure-Preserving Deep Learning with Calibrated Posterior Uncertainty under Incomplete Physics and Unknown Noise ​
Author: Minhyeok Ko, Abdollah Shafieezadeh
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, eess.SP, math.NA
arXiv:2608.04201v2 Announce Type: replace Abstract: Nonlinear state estimation requires sequentially fusing model-based predictions with noisy measurements. Under imperfect dynamics and unknown, time-varying noise statistics, this fusion can degrade in both accuracy and statistical consistency. Exis...
226. Continual Learning in Transition ​
Author: Zhiyan Hou, Dan Zhang, Tao Feng, Liyuan Wang, Wei Li, Xiangzhao Hao, Hongyan An, Junfeng Fang, Haokai Ma, Zhaohui Xu, Xinyu Tang, Haiyun Guo, Jinqiao Wang, Tat-Seng Chua
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.06216v2 Announce Type: replace Abstract: Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g., training strategies, architectural designs, and weight adaptation. However, emerging paradigms are...
227. ED-CSP: Crystal Structure Prediction from Electron Diffraction ​
Author: Germain Poloudenny, Ya"el Fr'egier, Arnaud Demorti`ere
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2608.06448v2 Announce Type: replace Abstract: Recovering a periodic 3D crystal structure from sparse, unindexed electron diffraction (ED) observations is a challenging generative inverse problem. Existing ED-based learning methods mainly predict crystallographic labels, reconstruct structures ...
228. Causal State-Space Model for Causal Inference: Estimating Longitudinal Individual Treatment Effects ​
Author: Abisoye Abidakun, Mingjun Zhong, Georgios Leontidis
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2608.08288v2 Announce Type: replace Abstract: Estimating counterfactual outcomes over time from longitudinal observational data is central to clinical decision support. Existing methods rely on domain confusion -- adversarial training that renders representations invariant to treatment assignm...
229. CurveFP: Co-Designing Numerical Representation and Product Arithmetic for Language Models ​
Author: Ye Qiao
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.10010v2 Announce Type: replace Abstract: Low-precision formats usually optimize scalar fidelity while inheriting conventional product arithmetic. We introduce CurveFP, a block-scaled family that distributes magnitudes across interleaved logarithmic curves. Uniform curve indices make every...
230. Do Judges Behave Like Algorithms? ​
Author: Riya Manchanda, Eric Chen, Chloe Zhu, Cynthia Rudin, Brandon Garrett, Songman Kang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2608.10400v2 Announce Type: replace Abstract: What if judges already behave like algorithms? As artificial intelligence and algorithms are deployed in many settings, including the judicial system, many have debated whether judges should be allowed to rely on them. Instead, we ask whether judge...
231. Optimize Cheap, Deploy Strong: Cost-Aware Cross-Tier Transfer for Evolutionary Optimization ​
Author: Tal Oved, Roi Pony, Oshri Naparstek, Udi barzelay
Published: 8/13/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.NE
arXiv:2608.10694v2 Announce Type: replace Abstract: Evolutionary optimization of LLM prompts and agentic programs (e.g., GEPA) is dominated by fitness evaluation: scoring each candidate runs an answering LLM over a validation set, so the evaluator's price tier dictates total search cost. We restruct...
232. Soft-Attention Improves Skin Cancer Classification Performance ​
Author: Soumyya Kanti Datta, Seyed Mohammad Abuzar Hashemi, Sargur N. Srihari, Mingchen Gao
Published: 8/13/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG
arXiv:2105.03358v4 Announce Type: replace-cross Abstract: In clinical applications, neural networks must focus on and highlight the most important parts of an input image. Soft-Attention mechanism enables a neural network toachieve this goal. This paper investigates the effectiveness of Soft-Attenti...
233. Heterogeneous transfer learning for high-dimensional regression with feature mismatch ​
Author: Jae Ho Chang, Massimiliano Russo, Subhadeep Paul
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2412.18081v3 Announce Type: replace-cross Abstract: We study Heterogeneous Transfer Learning (HTL) for high-dimensional regression with differing feature sets. Such feature mismatch arises when some variables available in a data-rich source domain are unavailable in a data-poor target domain. ...
234. A Variational Analysis of Kernel Learning with Learnable Linear Transformations ​
Author: Yang Li, Feng Ruan
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.CA, math.FA, math.OC
arXiv:2502.11665v3 Announce Type: replace-cross Abstract: The classical kernel ridge regression problem aims to find the best fit for the output $Y$ as a function of the input data $X\in \mathbb{R}^d$, with a fixed choice of regularization term imposed by a given choice of a reproducing kernel Hilbe...
235. SteeringSafety: Benchmarking Representation Steering in LLMs Across Safety Perspectives ​
Author: Vincent Siu, Nicholas Crispino, David Park, Nathan W. Henry, Zhun Wang, Yang Liu, Dawn Song, Chenguang Wang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2509.13450v3 Announce Type: replace-cross Abstract: We introduce SteeringSafety, a benchmark for evaluating representation steering methods across nine safety perspectives spanning 18 datasets. While prior work highlights the general capabilities of representation steering, we focus on safety ...
236. End-to-end Differentiable Calibration and Reconstruction for Optical Particle Detectors ​
Author: Omar Alterkait, C'esar Jes'us-Valls, Ryo Matsumoto, Patrick de Perio, Kazuhiro Terao
Published: 8/13/2026, 4:00:00 AM
Categories: hep-ex, cs.LG, physics.ins-det
arXiv:2602.24129v3 Announce Type: replace-cross Abstract: Large-scale homogeneous detectors with optical readouts are widely used in particle detection, with Cherenkov and scintillator neutrino detectors as prominent examples. Analyses in experimental physics rely on high-fidelity simulators to tran...
237. Post-Training with Policy Gradients: Optimality and the Base Model Barrier ​
Author: Alireza Mousavi-Hosseini, Murat A. Erdogdu
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2603.06957v2 Announce Type: replace-cross Abstract: We study post-training linear autoregressive models with outcome and process rewards. Given a context $\boldsymbol{x}$, the model must predict the response $\boldsymbol{y} \in Y^N$, a sequence of length $N$ that satisfies a $\gamma$ margin co...
238. LLM Router: Rethinking Routing with Prefill Activations ​
Author: Tanay Varshney, Annie Surla, Michelle Xu, Gomathy Venkata Krishnan, Maximilian Jeblick, David Austin, Neal Vaidya, Davide Onofrio
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2603.20895v3 Announce Type: replace-cross Abstract: Existing routers rely on semantic query features or handcrafted features, which often fail to capture model-specific failures or intrinsic task difficulty. We instead route using internal LLM activations, specifically the residual stream. Our...
239. CGRL: Causal-Guided Representation Learning for Node-Level Out-of-Distribution Generalization ​
Author: Bowen Lu, Lianqiang Yang, Teng Li, Kun Zhang
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2603.24304v3 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) deliver strong performance on graph tasks, but their accuracy drops significantly under out-of-distribution (OOD) scenarios. Under distribution shifts, GNNs often fit environmental noise and spurious correlations ...
240. Trust Region Constrained Bayesian Optimization with Penalized Constraint Handling ​
Author: Raju Chowdhury, Tanmay Sen, Biswabrata Pradhan
Published: 8/13/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2603.24567v2 Announce Type: replace-cross Abstract: Constrained optimization in high-dimensional black-box settings is difficult due to expensive evaluations, the lack of gradient information, and complex feasibility regions. In this work, we propose a Bayesian optimization method that combine...
241. Diffusion-Guided Cooperative Policy Learning for Target Tracking Based on Underwater Mobile Agent Networks ​
Author: Jiaao Ma, Chuan Lin, Guangjie Han, Shengchao Zhu, Zhenyu Wang, Chen An
Published: 8/13/2026, 4:00:00 AM
Categories: cs.NI, cs.LG
arXiv:2603.29426v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) provides a promising solution for cooperative target tracking in networks of autonomous underwater vehicles (AUVs). However, existing methods still face three major challenges: 1) policy non-stationar...
242. On Data-Driven Koopman Representations of Nonlinear Delay Differential Equations ​
Author: Santosh Mohan Rajkumar, Dibyasri Barman, Kumar Vikram Singh, Debdipta Goswami
Published: 8/13/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.DS
arXiv:2604.03086v2 Announce Type: replace-cross Abstract: This work establishes a rigorous bridge between infinite-dimensional delay dynamics and finite-dimensional Koopman learning, with explicit and interpretable error guarantees. While Koopman analysis is well-developed for ordinary differential ...
243. ReCodeAgent: A Multi-agent Workflow for Language-Agnostic Translation and Validation of Large-Scale Repositories ​
Author: Ali Reza Ibrahimzada, Brandon Paulsen, Daniel Kroening, Reyhaneh Jabbarvand
Published: 8/13/2026, 4:00:00 AM
Categories: cs.SE, cs.LG
arXiv:2604.07341v3 Announce Type: replace-cross Abstract: Most repository-level code translation and validation techniques have been evaluated on a single source-target programming language (PL) pair, owing to the complex engineering effort required to adapt new PL pairs. Programming agents can enab...
244. TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning ​
Author: Matthew M. Hong, Jesse Zhang, Anusha Nagabandi, Abhishek Gupta
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2605.12236v2 Announce Type: replace-cross Abstract: Fine-tuning pre-trained robot policies with reinforcement learning (RL) often inherits the bottlenecks introduced by pre-training with behavioral cloning (BC), which produces narrow action distributions that lack the coverage necessary for do...
245. Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models ​
Author: Qinwu Xu, Yifan Jiang, Haoyu Ren
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG
arXiv:2605.16409v3 Announce Type: replace-cross Abstract: Optical character recognition (OCR) and multilingual scene-text understanding remain challenging for multimodal large language models (MLLMs), particularly in real-world images containing small or degraded text, cluttered layouts, occlusion, ...
246. Large language models reorganize representational geometry during in-context learning ​
Author: Hua-Dong Xiong, Li Ji-An, Robert C. Wilson, Kwonjoon Lee, Xue-Xin Wei
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, q-bio.NC
arXiv:2605.28854v3 Announce Type: replace-cross Abstract: Large language models (LLMs) show remarkable flexibility in adapting to novel tasks without parameter updates, a capacity known as in-context learning (ICL). Prior work has sought to understand ICL by studying the circuits, algorithms, and re...
247. Moxia: A Trust-First Neuro-Symbolic Execution Architecture for Self-Explaining Mathematical Reasoning ​
Author: Alessio Bruno
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2606.00671v3 Announce Type: replace-cross Abstract: We present Moxia (formerly AXIOM), a trust-first neuro-symbolic architecture for self-explaining mathematical reasoning over natural-language input. Its language model is strictly a canonicalizer: it rewrites informal problem text into a narr...
248. Ranking vs. Assignment: The Metric Mismatch in Multi-View Object Association ​
Author: Matvei Shelukhan, Timur Mamedov, Aleksandr Chukhrov, Karina Kvanchiani
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2606.02022v2 Announce Type: replace-cross Abstract: Multi-view object association is an important computer vision problem that underlies many multi-camera perception tasks. While this task is naturally formulated as a constrained one-to-one matching problem, recent works heavily rely on pairwi...
249. RedditPersona: A Modular Framework for Community-Conditioned LLM Adaptation from Reddit ​
Author: Amirhossein Ghaffari, Ali Goodarzi, Huong Nguyen, Simo Hosio, Lauri Lov'en, Ekaterina Gilman
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.SI
arXiv:2606.06027v2 Announce Type: replace-cross Abstract: Community-conditioned language model adaptation needs choices about data collection, community definition, and evaluation that are currently made independently in each study, making it hard to compare assumptions or reuse artifacts. We presen...
250. FACTR 2: Learning External Force Sensing for Commodity Robot Arms Improves Policy Learning ​
Author: Steven Oh, Jason Jingzhou Liu, Tony Tao, Philip Han, Kenneth Shaw, Satoshi Funabashi, Ruslan Salakhutdinov, Deepak Pathak
Published: 8/13/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, cs.SY, eess.SY
arXiv:2606.12406v2 Announce Type: replace-cross Abstract: Contact-rich manipulation requires force sensitivity, but many robot arms lack dedicated force sensors due to their high cost. We present Neural External Torque Estimation (NEXT), a data-driven method that estimates external joint torques wit...
251. MMLA: How Memory Lets the Past Shape the Future ​
Author: Junyi Zou, Avrova Donz
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2606.28876v3 Announce Type: replace-cross Abstract: Proposal. Long context can replay history, but it does not decide which completed observations deserve authority. MMLA formalizes a bounded resident memory between transient context and slow weight updates. A completed local segment is eventi...
252. Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop ​
Author: Chenmu Zhang, Boris I. Yakobson
Published: 8/13/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.AI, cs.LG
arXiv:2606.29717v2 Announce Type: replace-cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has produced standard public benchmarks and many published machine-learning models for the task (D...
253. WBMM: Windowed Batch Matrix Multiplication for Efficient Large Receptive Field Convolution ​
Author: Wan Song, Wei Zhou, Rui Wang, Jun Yu, Toru Kurihara, Jiajia Xu, Shu Zhan
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.02097v2 Announce Type: replace-cross Abstract: Large kernel depthwise convolutions achieve strong performance but suffer from significant degradation as kernel size grows due to irregular memory access from gather-based computation; while Large Kernel Acceleration (LKA) helps on small fea...
254. OrderMoE: An expert similarity driven distributed edge MoE inference ​
Author: Xin Yuan, Ning Li, Quan Chen, Wenchao Xu, Song Guo
Published: 8/13/2026, 4:00:00 AM
Categories: cs.NI, cs.DC, cs.LG
arXiv:2607.17154v2 Announce Type: replace-cross Abstract: Although mixture-of-experts, MoE, models have been increasingly adopted to scale large language models with moderate computation cost, it remains challenging to deploy MoE inference over resource-constrained and bandwidth-limited edge infrast...
255. PathRIR: Physics-Guided Acoustic Path Selection and Late-Tail Compensation for Fast Room Impulse Response Simulation ​
Author: Shaoheng Xu, Chunyi Sun, Jihui Zhang, Amy Bastine, Prasanga N. Samarasinghe, Thushara D. Abhayapala
Published: 8/13/2026, 4:00:00 AM
Categories: eess.AS, cs.LG, cs.SD
arXiv:2607.23293v2 Announce Type: replace-cross Abstract: Image-source-method (ISM)-based room impulse response (RIR) simulation is a useful and physically interpretable tool for acoustic scene modeling, but full-order ISM becomes computationally expensive as the reflection order and room complexity...
256. DREvo: Distilling Recalibrated Historical Experience for Harness Self-Evolution ​
Author: Hanghui Guo, Weijie Shi, Zhangze Chen, Shengxiang Xu, Yishu Wang, Yimei Zhang, Wangze Ni, Jia Zhu, Shimin Di
Published: 8/13/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2607.26722v2 Announce Type: replace-cross Abstract: Harness plays a critical role in large language model agent performance, and building a high-performing harness requires substantial expert effort. Therefore, recent research has increasingly explored harness self-evolution, which iteratively...
257. A Constitution-Grid Instrument for Data-Efficient RL Alignment (C-Guard) ​
Author: Xianling Zhang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2608.00180v4 Announce Type: replace-cross Abstract: Conflicting objectives are general in RL alignment, and training on them data-efficiently is hard. Training a safety guard with RL means optimizing two objectives that conflict: catch real harm, and do not refuse benign prompts. Our finding i...
258. Rethinking Medical Landmark Localization with Prototype Learning-based Progressive Offset Correction ​
Author: Jingxian Xu, Yuhao Huang, Rusi Chen, Yanfeng Zhou, Dong Ni
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.09182v2 Announce Type: replace-cross Abstract: Accurate landmark localization in medical images is a fundamental step for quantitative clinical measurement and downstream analysis. Existing localization methods have advanced, among which multi-stage refinement is a superior solution. Alth...
259. Logit-Boundary Geometric Belief Interfaces and Sparse Sheaf-Enclave Protocols: A Self-Contained Substrate for Secure Network Electronic Health Record (EHR) Interoperability ​
Author: Alvin Spivey, Yu Huang
Published: 8/13/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG
arXiv:2608.10300v2 Announce Type: replace-cross Abstract: Electronic health-record interoperability is a boundary problem: legacy systems, generative models, terminology services, identity systems, and human reviewers may each expose rich internal states, while operational exchange requires a narrow...
260. When Do Anchor-Based Pointwise LLM Rerankers Help? Retriever Quality, Statistical Scope, and Anchor Design ​
Author: Utshab Kumar Ghosh, Shubham Chatterjee
Published: 8/13/2026, 4:00:00 AM
Categories: cs.IR, cs.LG
arXiv:2608.10528v2 Announce Type: replace-cross Abstract: Anchor-based pointwise LLM reranking scores each candidate against a shared reference passage to recover cross-document context at pointwise cost. We study when this actually helps, using GCCP/PAGC as a representative method. Our study is rep...
261. Beyond Fixed Luminance: Towards Panchromatic and Orthochromatic Image Colorization ​
Author: Swarnim Maheshwari, Syed Imam Ali, Vineeth N. Balasubramanian
Published: 8/13/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2608.10798v2 Announce Type: replace-cross Abstract: Most image colorization systems operate in $Lab$ space by predicting chroma ($ab$) while preserving an input-derived luminance channel ($L$). While effective on standard benchmarks, this fixed-luminance design restricts brightness changes and...