Skip to content

arXiv cs.LG - 2026-07-29 ​

245 items collected.


1. FinAbstain: Uncertainty-Calibrated Multimodal RAG for Selective Financial Forecasting ​

Author: Dorothy Torres, Wei Cheng, Henan Huang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.24875v1 Announce Type: new Abstract: Large language models (LLMs) can synthesize financial narratives but may express high confidence when evidence is sparse, stale, or contradictory. This failure is especially consequential in forecasting, where filings, news, prices, volume, and technic...

📖 Read original article


2. Human Preference aligned Tabular Similarity ​

Author: Frederik Hoppe, Astrid Franz, Marianne Michaelis, Lars Kleinemeier, Udo G"obel
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24880v1 Announce Type: new Abstract: Task-agnostic tabular embeddings are increasingly used for similarity search in real-world business systems such as Product Lifecycle Management (PLM). However, leading embedding approaches are optimized primarily for prediction tasks - not for produci...

📖 Read original article


3. Behavior-Driven Explainability ​

Author: Caroline Dominik, Rolf Drechsler
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.SE

arXiv:2607.24881v1 Announce Type: new Abstract: As system complexity has vastly increased, it has become significantly more challenging for a single person or a team to fully understand all aspects of an entire system. Particularly, this holds when considering all the different stages of a system's ...

📖 Read original article


4. Eliminating Propagation Delay: Attention-Based Spatial-Temporal Fusion Graph Convolution Network for Traffic Flow Prediction ​

Author: Jinpeng Chen, Ziyu Yu, Tao Wang, Jun Ma, Hongbo Gao, Senzhang Wang, Zufeng Zhang, Kaimin Wei
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24885v1 Announce Type: new Abstract: Predicting traffic flow is crucial to optimizing transportation systems and improving urban mobility. Many graph convolution-based models have been proposed to extract spatial-temporal features and predict traffic flow. However, most focus on spatial-t...

📖 Read original article


5. Mechanisms of Width Scaling in Normalized Residual Networks: The Effective Alignment Dimension ​

Author: Jinhao Zhang, Zeyu Liu, Zicheng Yan, Yunquan Zhang, Guangming Tan, Fangming Liu, Daning Cheng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24887v1 Announce Type: new Abstract: Existing theories of neural-network width characterize asymptotic limits, but provide limited guidance on whether an expansion direction identified from finite training data remains beneficial on unseen data. We study this problem for function-preservi...

📖 Read original article


6. GAUGE: Grading Agent-Built Financial Models Without a Golden Answer ​

Author: Jiacheng Lu, Sinuo Wang, Wentao Zhao, Rui Sun, Cheng Hua, Tao Song, Hui Cai, Beidi Luan, Zhengze Wu, Lingjing Teng, Yijia He, Jing Li, Daxin Jiang, Zuo Bai, Haibing Guan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE

arXiv:2607.24889v1 Announce Type: new Abstract: Financial models combine public disclosures with analyst assumptions to produce forecasts and valuations. While some components can be checked mechanically, forecasts, discount rates, and target prices often admit multiple reasonable answers. Existing ...

📖 Read original article


7. LLM as Forecasting Planner: Training-Free Text Conditioning for Time-Series Foundation Models ​

Author: Huu Hiep Nguyen, Dung Nguyen, Minh Hoang Nguyen, Dai Do, Hung Le
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24892v1 Announce Type: new Abstract: Text-conditioned time-series forecasting predicts a series from both its numerical history and natural-language context, allowing forecasts to account for events and constraints that the past alone cannot reveal. This requires both reliable numerical f...

📖 Read original article


8. Inverse RL Helps Align AI by Imitating Humans ​

Author: Micha{\l} Wili'nski, Liu Leqi, Chirag Nagpal
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.24900v1 Announce Type: new Abstract: Language model alignment aims to make model behavior reliably reflect desirable properties such as helpfulness, safety, and instruction following. Current approaches typically use supervised fine-tuning on demonstrations or reinforcement learning with ...

📖 Read original article


9. Multiclass Classification without Labels via Posterior Simplex Geometry ​

Author: Rapha"el Bonnet-Guerrini, Johann Ioannou-Nikolaides, Troels Petersen, Vincenzo Piuri
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, astro-ph.GA, cs.AI, stat.ML

arXiv:2607.24943v1 Announce Type: new Abstract: In many classification problems, reliable instance-level labels are unavailable. However, it is often possible to construct weakly enriched unlabeled samples: datasets selected by different cuts, sources, populations, or experimental conditions that ch...

📖 Read original article


10. Stable FP4 Training via Transposition-Invariant Block Quantization ​

Author: Mehdi Rahimifar, Amin Darabi, Mehran Taghian Jazi, Xing Huang, Yao Wang, Zhijun Tu, Yufei Cui, Yunke Peng, Hongliang Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.24953v1 Announce Type: new Abstract: Reducing training precision is a key lever for improving the e ciency of large language model (LLM) training, but pushing beyond FP8 to 4-bit oating point (FP4) remains challenging due to instability during optimization. We identify a fundamental sourc...

📖 Read original article


11. Generative Distributionally Robust Optimization ​

Author: Ziwei Zhang, Jonathan Yu-Meng Li, Zhihao Jin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC, stat.ML

arXiv:2607.24983v1 Announce Type: new Abstract: Generative models are increasingly adopted in distributionally robust optimization (DRO), but existing approaches trade off model compatibility and adversarial structure: methods that accept arbitrary samplers do not restrict worst-case laws to a gener...

📖 Read original article


12. Calibrated Partial Resets: Preventing Policy Collapse in Continual Reinforcement Learning ​

Author: Luc McCutcheon, Evangelos Chatzaroulas, Saber Fallah
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO

arXiv:2607.24996v1 Announce Type: new Abstract: Neural networks are hindered by accumulating dormant neurons and loss of expressivity throughout training, particularly in non-stationary data settings, such as continual supervised and reinforcement learning. Recently, neuron resets have been used to ...

📖 Read original article


13. Conformal Cascade: Distribution-Free Accuracy Guarantees for Multi-Tier LLM Inference ​

Author: Yifan Dou, Shikan Fang, Shibo Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25018v1 Announce Type: new Abstract: Large language model (LLM) cascades reduce inference cost by routing easy queries to a small model and deferring hard queries to a larger one. Production cascades govern this deferral through a confidence threshold, but LLM confidence scores are miscal...

📖 Read original article


14. Lantern: Conflict-Aware Gradient Blending for Physics-Guided Diffusion Models in Calorimeter Simulation ​

Author: Farzana Yasmin Ahmad, Vanamala Venkataswamy, Geoffrey Fox
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25060v1 Announce Type: new Abstract: Monte Carlo simulation of calorimeter showers is a principal bottleneck for the High-Luminosity LHC, and diffusion models have emerged as fast, high-fidelity surrogates. Their denoising objective is purely statistical, however: a model can minimize it ...

📖 Read original article


15. Score-Based Stabilization for Time-Dependent Problems ​

Author: Eshed Gal, Eldad Haber, Uri Ascher
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2607.25119v1 Announce Type: new Abstract: We propose a score-based stabilization framework for numerical simulation of partial differential equations, in which a learned score model defines a stabilization operator applied to provisional numerical updates. This operator augments standard time-...

📖 Read original article


16. Semantic Space Search Trajectory Networks ​

Author: Julian Agudelo, Alberto Tonda, Gabriela Ochoa, Vincent Guigue, Cristina Manfredotti, Evelyne Lutton
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25122v1 Announce Type: new Abstract: Search Trajectory Networks (STNs) are a graph-based tool for visualizing and characterizing the behavior of optimization algorithms. STNs' reliance on discretization of the search space has largely confined them to low-dimensional or combinatorial sett...

📖 Read original article


17. Endpoint Replay: Compressing the Recency Buffer in Deep Reinforcement Learning ​

Author: Parham Mohammad Panahi, Armin Ashrafi, Haoyu Du, Andrew Patterson, Martha White, Adam White
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25123v1 Announce Type: new Abstract: Experience replay remains one of the most practical and useful algorithmic tools in the deep reinforcement learning (DRL) toolbox. Aside from the limited success of prioritized replay and specialized approaches for large asynchronous systems, most DRL ...

📖 Read original article


18. Interpretable GOHR Agents via Sparse Autoencoders ​

Author: Shiwei Tan, Yusong Zhao, Weiyi Qin, Wentian Wang, Jacob Feldman, Lazaros K. Gallos, Paul B. Kantor, Vladimir Menkov, Hao Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25132v2 Announce Type: new Abstract: A central challenge in interpreting learned decision-making systems is to determine whether their internal representations contain concepts that help explain their behavior. We report interpretability experiments for a tokenized autoregressive Transfor...

📖 Read original article


19. Physics-Informed CNN-LSTM for Street-Scale Urban Flood Prediction: Reconciling Aggregate Accuracy and Street-Level Plausibility ​

Author: Luc DCosta, Yidi Wang, Jonathan L. Goodall, Rohan Chandra
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25148v1 Announce Type: new Abstract: Deep learning surrogate models trained with mean-squared-error loss produce statistically accurate but physically unconstrained flood predictions: water may flow uphill, appear spontaneously, or smooth over street-level corridors. We develop a physics-...

📖 Read original article


20. Accurate structural modeling of chemically diverse molecular interfaces with Vilya-2 ​

Author: Vilya Research, :, Pascal Sturmfels, Naozumi Hiranuma, Milad Salem, Benjamin D. Sellers, Stephen Rettie, CJ San Felipe, Chase A. P. Wood, Jeffrey K. Holden, Adam P. Moyer, Patrick J. Salveson, Ivan Anishchanka
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25156v1 Announce Type: new Abstract: Structure-prediction networks built on co-evolutionary statistics have transformed protein-based drug discovery, yet their accuracy does not extend to peptide therapeutics--an increasingly important modality defined by non-canonical residues, macrocycl...

📖 Read original article


21. CondPSE: A Polynomial-Filtered Structural Encoder with Conditional Modulation for Graphs ​

Author: Woohyun Lee, Hogun Park
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25169v1 Announce Type: new Abstract: Message-passing graph neural networks are bounded by the 1-WL test and can miss topological structure that distinguishes non-isomorphic graphs. Positional and structural encodings (PSE) inject such topology-derived signals, and learned PSE encoders suc...

📖 Read original article


22. Rethinking CD: A Reproducibility Study and Extension on the Ineffectiveness of Contrastive Decoding at Mitigating Object Hallucinations in MLLMs ​

Author: Arnav Bendre, Guneesh Gupta, Kavish Grover, Chayan Aggarwal, Shreyansh Modi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25196v1 Announce Type: new Abstract: Contrastive decoding (CD) has been proposed as a training-free strategy for mitigating object hallucinations in multimodal large language models (MLLMs), with reported gains on benchmarks such as POPE. However, recent work has questioned whether these ...

📖 Read original article


23. Algorithmic Separation between Constant-Depth and Logarithmic-Depth Neural Networks ​

Author: Yunwei Ren, Zihao Wang, Jason D. Lee
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.25200v1 Announce Type: new Abstract: Despite the empirical advantages of deep networks over shallow ones, theoretical depth separations largely concern approximation power, while algorithmic results are mostly limited to comparisons between two- and three-layer networks. In this work, we ...

📖 Read original article


24. A Unified Algorithmic Framework for Hybrid Reinforcement Learning in Tabular MDPs with Shifted Transition Dynamics ​

Author: Zheshun Wu, Renjie Zheng, Jinhang Zuo, Zenglin Xu, Fang Kong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25207v1 Announce Type: new Abstract: This paper investigates a hybrid reinforcement learning setting in tabular Markov Decision Processes (MDPs), where an agent aims to learn an optimal policy by combining online interactions with a target environment and offline data from a source enviro...

📖 Read original article


25. Neurai-VN Benchmark: Standardized Machine Learning Models for Multimodal Digital Phenotyping in Mental Health Classification ​

Author: Quoc-Cuong Pham, Hoang-Thuy-Duong Vu, Thi-Thanh-Huong Ha, Huy-Hieu Pham
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25232v1 Announce Type: new Abstract: Digital phenotyping (DP) using smartphones and wearable devices has shown considerable potential for mental health monitoring. However, progress remains difficult to evaluate due to heterogeneous datasets, inconsistent preprocessing pipelines. In this ...

📖 Read original article


26. Beyond Single-Episode Optimization: Sliding-Window Aware Generative Auto-Bidding for Long-Term Advertising Effectiveness ​

Author: Binglin Wu, Chuan Yue, Yingyi Zhang, Xianneng Li, Ruyue Deng, Weiru Zhang, Xiaoyi Zeng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25233v1 Announce Type: new Abstract: Auto-bidding systems optimize bids to maximize value under efficiency constraints such as Cost-Per-Action (CPA). Existing methods treat each day as an independent episode. However, many advertisers produce value so sparsely that per-day efficiency rati...

📖 Read original article


27. Bridging Compute- and Data-Optimal Pretraining ​

Author: Tian Qin, Kimia Hamidieh, David Alvarez-Melis
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.PF

arXiv:2607.25271v1 Announce Type: new Abstract: Classical compute-optimal scaling laws assume an unbounded supply of fresh pretraining data, yet pretraining is increasingly entering a regime in which compute grows faster than the availability of high-quality data. We propose Compute-Data (CD) scalin...

📖 Read original article


28. HeAD-CP: Heterophily-Aware Diffused Conformal Prediction Sets for Graph Neural Networks ​

Author: Phan Binh Nguyen Lam, Nguyen Thai Anh
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25273v1 Announce Type: new Abstract: Conformal prediction (CP) provides distribution-free uncertainty quantification, and its extension to graphs is an active research direction. Diffused Adaptive Prediction Sets (DAPS) is a widely used graph-aware diffusion baseline, propagating Adaptive...

📖 Read original article


29. When Does Deep Representation Learning Help Single-Cell Clustering? A Sensitivity-Aware Diagnostic Benchmark for Biomedical AI Pipelines ​

Author: Nguyen Thanh Phong, Truong Viet Vu, Nguyen Ha Thu, Tran An Ky, Tran Hoang Thong, Le Pham Thuy Hien, Nguyen Thai Anh
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, q-bio.GN

arXiv:2607.25288v1 Announce Type: new Abstract: Single-cell ribonucleic acid sequencing (scRNA-seq) is a foundational technology for precision-medicine workflows that contribute to United Nations Sustainable Development Goal 3 on Good Health and Well-being, and unsupervised clustering is the analyti...

📖 Read original article


30. AMRD: Adaptive Multi-Teacher Relational Distillation for Lightweight Speech Emotion Recognition ​

Author: Yuqi Li, Yi-Cheng Lin, Xianglong Wang, Kuo Yang, Xiaoqin Feng, Yixuan Wang, Huiran Duan, Yingli Tian
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25289v1 Announce Type: new Abstract: On-device speech emotion recognition (SER) is critical for real-time applications, yet large self-supervised models that excel at SER are too costly for edge devices. Multi-teacher knowledge distillation can compress them into a lightweight student, bu...

📖 Read original article


31. Breaking the Periodicity Assumption: Robust Tensorial Multi-View Clustering via Graph-Spectral Low-Rank Learning ​

Author: Jintian Ji, Xingsu Li, Songhe Feng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25295v1 Announce Type: new Abstract: Tensorial multi-view clustering (TMC) has achieved strong performance due to its ability to capture high-order correlations across multiple views. Most existing t-SVD-based TMC frameworks apply the Fast Fourier Transform (FFT) along the sample mode to ...

📖 Read original article


32. Zhinv: Real-time hub-height wind field reconstruction using only local sparse observations ​

Author: Zongwei Zhang, Chin Chun Ooi, Lianlei Lin, Sheng Gao, Tiantian He, Yew Soon Ong, Junkai Wang, Hangyi Yu, Jiaqi Zhang, Hanqing Zhao, Yu Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25298v1 Announce Type: new Abstract: The high proportion of wind power connected to the grid places higher demands on fine-grained knowledge of regional wind fields. Since the wind information directly obtainable in actual operations is mostly sparse, discrete, and irregularly distributed...

📖 Read original article


33. Retraction-Free Optimization over the Stiefel Manifold for the LoRA Fine-Tuning ​

Author: Yuan Zhang, Jiang Hu, Zhijian Lai, Lin Lin, Zaiwen Wen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25299v1 Announce Type: new Abstract: Optimization over the Stiefel manifold plays a significant role in various machine learning tasks. Existing methods either use the retraction operators, requiring costly orthonormalization for large-scale matrices, or employ landing methods that rely o...

📖 Read original article


34. Guiding Posterior Exploration with Optimizer-Derived Geometry ​

Author: Moritz Schlager, Emanuel Sommer, Thomas M"ollenhoff, David R"ugamer
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25312v1 Announce Type: new Abstract: Sampling-based methods offer a principled approach to uncertainty quantification in Bayesian neural networks. Their practical use, however, is often challenged by the computational cost of exploring high-dimensional and multimodal posterior distributio...

📖 Read original article


35. Explainable AI for Chronic Kidney Disease Prediction Using Simulated Federated Learning ​

Author: Md Zahid Hasan Ontor, Md Al Amin, Anik Dev Nath, Bikash Kumar Paul
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25348v1 Announce Type: new Abstract: Chronic Kidney Disease (CKD), characterized by the gradual loss of kidney function, remains a significant public health challenge. Early detection is crucial for preventing severe complications and enhancing patient outcomes. In this study, Federated L...

📖 Read original article


36. Raven: High-Recall Sequence Modeling with Sparse Memory Routing ​

Author: Arshia Afzal, Aviv Bick, Eric P. Xing, Volkan Cevher, Albert Gu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25357v1 Announce Type: new Abstract: Long-context recall in linear-time sequence models highlights a tradeoff in how they write to memory. State-based linear models, such as state-space models (SSMs) and linear Transformers, write densely, updating the entire state for each newly arrived ...

📖 Read original article


37. Rethinking Likelihood distributions: Student's t Likelihood Boosts Bayesian Neural Network Performance ​

Author: Pei-Hsuan Hsia, Lars H. Heyen, Arvid Weyrauch, Markus Goetz, Achim Streit, Sebastian Krumscheid, Charlotte Debus
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25376v1 Announce Type: new Abstract: In Bayesian neural networks (BNNs), variational inference is a widely adopted framework for modeling uncertainty in a distributional way, with the evidence lower bound (ELBO) serving as the standard objective function. Several distributions contribute ...

📖 Read original article


38. Learned, Relied Upon, or Necessary? Separating Checkpoint Dependence from Task-Level Value in Sheaf GNNs ​

Author: Yi Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25387v1 Announce Type: new Abstract: Learned restriction maps in sheaf graph neural networks are often treated as proof that the model has discovered useful edge geometry. That conclusion does not follow from parameter movement or from a post-hoc ablation: both can show how one checkpoint...

📖 Read original article


39. TWICE: Two-Clock, Two-Window Learning for Long-Horizon Conversion Prediction in Online Advertising ​

Author: Kaiyuan Li, Kun Wang, Zhongbo Wang, Teng Sha, Ming Yan, Yanhua Cheng, Xialong Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2607.25404v1 Announce Type: new Abstract: Long-horizon conversion prediction under delayed feedback creates a two-clock, two-window learning problem in online advertising. A short base observation window releases recent clicks on the click clock before their outcomes mature, whereas conversion...

📖 Read original article


40. SPARC Segmentation to Prediction via Affine Regression and Counterfactuals ​

Author: Shivani, Subhayan Roy
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25413v1 Announce Type: new Abstract: Transaction propensity prediction in B2B e commerce presents unique challenges distinct from B2C contexts, primarily due to the heterogeneous procurement behaviors of organizational entities, which violate SMOTE's implicit assumption of within class fe...

📖 Read original article


41. PIcsC: Partitioning-Induced Covariate Shift Correction ​

Author: Behraj Khan, Behroz Mirza, Syed Ahmad Chan Bukhari, Tahir Qasim Syed
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25441v1 Announce Type: new Abstract: Covariate shift across training-data partitions biases model selection and parameter estimation in cross-validation, lifelong learning, and federated learning. We propose \textit{Partition-Induced Covariate-shift Correction} (\texttt{PIcsC}), a Fisher ...

📖 Read original article


42. Bits and Memories: Measuring Verbatim Extraction Across LLM Quantization ​

Author: Akshay Sasi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2607.25451v1 Announce Type: new Abstract: Language models are almost always quantized before they are deployed, and a growing line of work asks whether quantization also lowers their privacy risk. That work measures privacy almost entirely with membership inference. We think this is the wrong ...

📖 Read original article


43. Emergent Latent-State Computation under Stochastic Volatility ​

Author: Xiaoyu Huang, Lulu Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-fin.ST

arXiv:2607.25459v1 Announce Type: new Abstract: Mechanistic interpretability has largely focused on language models and deterministic toy tasks. Much less is known about how sequence models internally represent latent stochastic dynamics under noisy, partially observed observations. We study this qu...

📖 Read original article


44. Data-Dependent Regret and Polyak Corrections for Constrained Online Convex Optimization ​

Author: Wentao Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25480v1 Announce Type: new Abstract: Constrained online convex optimization requires minimizing regret against adversarial convex costs while satisfying a convex constraint at every round, as needed in safety-critical applications. A computationally efficient method combines online gradie...

📖 Read original article


45. Quantum Speedups for Stochastic Optimization with Heavy-Tailed Noise ​

Author: Bin Luo, Chengchang Liu, Jonathan Allcock, Shengyu Zhang, John C. S. Lui
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25492v1 Announce Type: new Abstract: We study stochastic optimization with heavy-tailed gradient noise. We first propose a novel quantum mean estimator for multivariate heavy-tailed random variables that achieves lower query complexity than optimal classical estimators in the low-dimensio...

📖 Read original article


46. Anti-Backdoor Coreset Selection via Cumulative Entropy ​

Author: Qi Zhao, Christian Wressnegger
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2607.25502v1 Announce Type: new Abstract: Recent training-time defenses against neural backdoors isolate a benign subset from poisoned training data, to learn a backdoor-free model from it. In this paper, we formulate this defense strategy as a coreset selection problem, giving rise to so-call...

📖 Read original article


47. AMPBench-MT: A Homology-Controlled Benchmark for Antimicrobial Peptide Potency, Spectrum, and Safety Prediction ​

Author: Ziheng Zhou, Huiyu Luo, Xiaohu Zhu, Nan Wang, Xuebiao Qin, Chaoyan Zhang, Jun Yan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM

arXiv:2607.25518v1 Announce Type: new Abstract: Computational AMP discovery is often evaluated through AMP/non-AMP recognition, yet follow-up decisions depend on assay-derived evidence such as target-species potency, hemolysis, toxicity, and selectivity. Existing AMP and peptide benchmarks cover bin...

📖 Read original article


48. Multi-Scale Structural Features for Continual, Comprehensible Visual Recognition in a Developmental Learning Framework ​

Author: Zeki Doruk Erden
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.25531v1 Announce Type: new Abstract: Contemporary machine learning struggles to learn continually, reuse prior knowledge, and expose a comprehensible internal structure. A recently proposed developmental, gradient-free learning framework addresses these limitations by learning a discrete,...

📖 Read original article


49. Mind the Missing Split: Resolving Feature Heterogeneity in Swarm Learning with Random Forests ​

Author: Mohammad Tajabadi, Dominik Heider
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25538v1 Announce Type: new Abstract: Swarm Learning is a decentralized collaborative learning mechanism that allows multiple organizations to train a shared model without central coordination or direct data sharing. In typical horizontal Swarm Learning, datasets across sites are usually a...

📖 Read original article


50. OrthKD: Extracting Generalized Clinical Knowledge from Heterogeneous Teachers for Lightweight Deployment ​

Author: Yi Xu, Cheng Chen, Mufan Cao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.25545v1 Announce Type: new Abstract: Deploying diabetic retinopathy (DR) screening models in primary care requires edge-efficient systems that remain accurate, safe, and reliable under domain shift. Multi-teacher knowledge distillation (KD) is a natural compression strategy, but existing ...

📖 Read original article


51. Physics-Informed Broad Learning System: An Efficient Backpropagation-Free Framework for Solving Partial Differential Equations ​

Author: Pinki Khatun, M. Sajid, Abhinav Jha, M. Tanveer
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25608v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful paradigm for solving partial differential equations (PDEs) by embedding governing physical laws into deep neural networks. However, their reliance on computationally expensive gradient...

📖 Read original article


52. Contrastive Representation Learning of Longitudinal Disease Trajectories on Temporal Graphs ​

Author: Bastian Pfeifer
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.QM

arXiv:2607.25609v1 Announce Type: new Abstract: Understanding disease trajectories from longitudinal clinical data remains challenging due to complex temporal dynamics and heterogeneous patient cohorts. Here, we present a contrastive representation learning framework that models multivariate disease...

📖 Read original article


53. MemSFT: Mitigating Alignment Tax with an External Parametric Memory ​

Author: Jiarui Wang, Xiang Shi, Jiaqi Cao, Rubin Wei, Xiquan Wang, Hao Sun, Jingzhi Wang, Zhiqi Yang, Qipeng Guo, Bowen Zhou, Zhouhan Lin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.25614v1 Announce Type: new Abstract: Adapting Large Language Models (LLMs) to specialized domains often incurs an alignment tax, as fine-tuning on domain-specific tasks can cause catastrophic forgetting and substantially degrade performance on general tasks. We propose MemSFT, which mitig...

📖 Read original article


54. Using Data-Derived Priors to Guide CNN Architecture Design for NIR Chemometrics ​

Author: D'ario Passos
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, physics.app-ph, physics.comp-ph

arXiv:2607.25636v1 Announce Type: new Abstract: Convolutional neural networks (CNN) for near-infrared (NIR) chemometrics are often designed using generic architectural rules, although spectral datasets differ in sampling, smoothness, redundancy, and sample size. We tested whether these properties ca...

📖 Read original article


55. Contextual Deconvolution for Variance-Stable Demand Sensing: Kernel-Modulated Operators in Promotional Retail ​

Author: Mohammad Forouhesh
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2607.25664v1 Announce Type: new Abstract: Machine learning demand forecasts optimize statistical accuracy yet leave excess operational volatility that inflates safety stock and amplifies the Bullwhip effect. We introduce \textbf{Contextual Deconvolution} (CD), a two-stage estimator that refram...

📖 Read original article


56. A Physics-Informed Neural Operator for Thermal Ranking of Low-Cost Wall Materials in Hot-Dry Climates ​

Author: Muhammad Akbar Khan, Fahim Raees, Ubaida Fatima
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, physics.comp-ph

arXiv:2607.25668v1 Announce Type: new Abstract: Identifying cost-effective indigenous building materials that minimise heat penetration through walls is critical for indoor thermal comfort in low-income rural housing in hot-dry climates, where summer temperatures routinely exceed 45 C. We present a ...

📖 Read original article


57. DynaBridge: Dynamic Summary-Guided Cross-Task Multimodal Fusion for DASS-Structured Mental Health Assessment ​

Author: Shiyu Teng, Haichen Yu, Jiaqing Liu, Hao Sun, Yu Song, Shurong Chai, Ruibo Hou, Lanfen Lin, Yen-Wei Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MM

arXiv:2607.25679v1 Announce Type: new Abstract: Multimodal behavioral analysis offers a scalable approach to assessing depression, anxiety, and stress, yet generic fusion models often ignore the psychometric structure of questionnaire labels. In DASS-21, risk labels are derived from ordered symptom ...

📖 Read original article


58. Rashomon Alignment ​

Author: Mois'es Santos, Peter van der Putten, Bernhard Pfahringer, Carlos Soares
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25680v1 Announce Type: new Abstract: We propose Rashomon Alignment (RA), a new measure to assess functional similarity between two models. Existing functional similarity measures are distributional, quantifying differences between outputs of models applied to real-world data. However, the...

📖 Read original article


59. From Deterministic to Generative Deep Learning for Urban Air Quality Reconstruction from Sparse Observations ​

Author: Abhishek A. Sabnis, Mihai Mitrea, Lya Lugon, Karine Sartelet, Marc Bocquet, Xiaoyuan Cheng, Shupeng Zhu, Sibo Cheng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25687v1 Announce Type: new Abstract: Full-field reconstruction of air pollution is essential for evaluating pollution exposure and supporting public health decision-making. However, the complex interactions among pollutants, hard-to-predict weather patterns, and limited monitoring station...

📖 Read original article


60. Tools Are Not Islands: Set-Level Tool Retrieval for LLM Agents via Query-Conditioned Hyperedge Prediction ​

Author: Xinyi Hong, Pinjun Dong, Xinyang Yu, Binyan Jiang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR

arXiv:2607.25718v2 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on invoking external tools to complete real-world tasks. Tool retrieval, which selects a small task-relevant subset from a library of thousands of tools before the agent acts, has therefore become a c...

📖 Read original article


61. Optimization with Dynamic Constraint Learning (DCL) ​

Author: Ezgi Oztekin, Figen Oztoprak, S. Ilker Birbil
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2607.25719v1 Announce Type: new Abstract: We propose Dynamic Constraint Learning (DCL), a data-driven framework for constrained optimization when constraint functions are unknown and cannot be queried during optimization. At each iteration, the method learns a local surrogate from nearby data ...

📖 Read original article


62. Detecting CSAM Text-to-Image LoRAs From Weights ​

Author: David Demitri Africa, Cate Heine, Nadine Staes-Polet, Kimberly Mai
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.CY

arXiv:2607.25750v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) fine-tuning has made it cheap and easy to customize open-weight image generation models for specific tasks, including the production of child sexual abuse material (CSAM). Existing moderation relies on metadata or generated o...

📖 Read original article


63. An Embarrassingly Simple Rule-based Visiting Circulation Approach to Trip Destination Prediction ​

Author: Eng-Shen Tu, Yong-Han Chen, En-Chao Liu, Hao-Yun Keng, Cheng-Te Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25751v1 Announce Type: new Abstract: In this paper, we propose the Rule-based Visiting Circulation (RVC) model in tackling the challenge in the IEEE Big Data Cup 2022: Trip Destination Prediction. Given trips containing travel information, personal attributes, origin zones, and their feat...

📖 Read original article


64. SpectONet: A Physics-Guided Spectral Deep Operator Network for Euler-Bernoulli Beam Dynamics ​

Author: Shivani Saini, Ramesh Kumar Vats, Arup Kumar Sahoo
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DS

arXiv:2607.25790v1 Announce Type: new Abstract: This paper proposes a novel physics-guided spectral deep operator network, termed SpectONet, for solving Euler-Bernoulli beam (EBB) vibration problems. The proposed framework integrates the operator-learning capability of DeepONet with physics-informed...

📖 Read original article


65. Prototype Adaptation for Zero-Shot sEMG Movement Classification ​

Author: Rui Liu, Benjamin Paassen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25826v1 Announce Type: new Abstract: Surface electromyography (sEMG) enables the control of prostheses, allowing upper-limb amputees to re-gain some hand function. Most current research focuses on recognizing basic movements for prosthesis control. However, in most daily activities, such ...

📖 Read original article


66. DRIFT: Direct-Recursive Intervention-Conditioned Forecasting of ICU Physiological Trajectories ​

Author: Weixin Liu, Juming Xiong, Congning Ni, Yanfan Zhu, Xingtao Lin, Bradley A. Malin, Zhijun Yin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2607.25864v1 Announce Type: new Abstract: Many time-series forecasts depend not only on prior observations but also on actions specified during the forecast period. In intensive care units (ICUs), future vital signs and laboratory values are influenced by treatments such as vasopressors. Howev...

📖 Read original article


67. A2TTA: Anchored-and-Agile Test-Time Adaptation for Evolving Traffic Sensor Networks ​

Author: Du Yin, Xiachong Lin, Yue Tan, Jinliang Deng, Estrid He, Hao Xue, Flora D. Salim
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25875v1 Announce Type: new Abstract: Traffic forecasting is important for efficient traffic management and route planning in smart cities. Existing traffic forecasting studies typically assume fixed sensor graphs, overlooking the continuous evolution of real-world traffic networks, e.g., ...

📖 Read original article


68. A Machine-Learning-Based Gas Lift Optimization Workflow for Unconventional Fields ​

Author: Sha (Sasha), Miao, Alexandra Vendetti, Logan Smart, Gunta Chomchalerm, Yang Chen, Christopher Frazier, Dustin Haralson, Jeremy Sorenson, Xiao Ma, Huafei Sun, Aaron Shinn, Haining Zheng, Xiao-Hui Wu, Peng Xu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SE

arXiv:2607.25885v1 Announce Type: new Abstract: In this paper, we present an automated data-driven workflow using Machine Learning (ML) for gas lift optimization in unconventional fields. This workflow integrates a ML model that accurately forecasts the Gas Lift Performance Curve, and a Bayesian Opt...

📖 Read original article


69. Minimizing Targeted Activations: Input-Only Suppression of Evaluation-Awareness Latents in Large Language Models ​

Author: Deepanshu Mody, Samarth Agarwal, Utkarsh Mittal, Dipesh Mahato
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.25907v1 Announce Type: new Abstract: Activation steering controls model behavior by editing internal activations at inference time. We study its input-side dual: optimizing a fluent prompt so that a chosen internal latent is driven toward zero, with no inference-time model access. Our tar...

📖 Read original article


70. Reinforcement Learning for Code Optimization ​

Author: Pierre Chambon, Kunhao Zheng, Juliette Decugis, Benoit Sagot, Gabriel Synnaeve
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25970v1 Announce Type: new Abstract: RL for code correctness is now established: have the model generate a program, run it against hidden test cases, and reward solutions that pass. Extending this to code optimization seems straightforward: just add execution time to the reward. But in pr...

📖 Read original article


71. Generator-Aligned Representation Interfaces for Diagnostic Soft Equivariance ​

Author: Weitao Li, Gong Cheng
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.25988v1 Announce Type: new Abstract: Exact-equivariant architectures typically encode prescribed group actions in specialized operators, which can complicate their reuse with generic backbones and across data modalities. We introduce the Generator-Aligned Representation Interface (GARI), ...

📖 Read original article


72. Empirical Evaluation of Out-Of-Distribution Performance of Tabular Foundation Models ​

Author: Malena Loza, David Chushig-Muzo, Eva Milara, Luis Bote-Curiel, Luis Estrada-Petrocelli, Felipe Grijalva
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.26000v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) have emerged as novel approaches for tabular predictive tasks, demonstrating competitive predictive performance to ensemble tree-based models. Most TFMs are trained and evaluated on independent and identically distribut...

📖 Read original article


73. Sharpness-Aware Minimization and Muon: Robustness under the Spectral Norm ​

Author: Wenzhi Zhong, Edward Milsom, Michael Murray
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.26001v1 Announce Type: new Abstract: Sharpness-Aware Minimization (SAM) aims to improve generalization by encouraging insensitivity to small, worst-case parameter perturbations. However, the notion of a "small" perturbation is inherently geometry-dependent: while existing SAM variants hav...

📖 Read original article


74. Reinformed Dreamer: An Asymmetric World Model Efficiently Trained through Latent Guidance ​

Author: Gaspard Lambrechts, Adrien Bolland, Daniel Ebi, Damien Ernst
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.26040v1 Announce Type: new Abstract: Much like humans benefit from guidance while learning, reinforcement learning algorithms may benefit from additional supervision beyond rewards. Leveraging additional information during training to learn better representations and behaviors has been th...

📖 Read original article


75. Re-thinking Mammography Transfer Learning: The Dataset-Informed Transfer Learning (DITL) Framework for Breast Cancer Screening and Lesion Diagnosis ​

Author: Adarsh Bhandary Panambur, Siming Bayer, Andreas Maier
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.26043v1 Announce Type: new Abstract: Enhancing classification performance in mammography remains a persistent challenge across both small curated datasets and large-scale clinical cohorts. Conventional transfer learning approaches often neglect dataset-specific characteristics, while rece...

📖 Read original article


76. Spend Experts Where You Are Unsure: Confidence-Adaptive Routing for Mixture-of-Experts LoRA ​

Author: Tom Saliencro, Rohan Desai, Priya Nair, Maya Lindqvist, Daniel Whitmore
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.26052v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) variants of Low-Rank Adaptation (LoRA) route every token to a fixed number of experts $k$. Tokens differ in how uncertain the model is about them, so a single k over-spends on easy tokens and under-serves hard ones. We observe ...

📖 Read original article


77. A Generalized Tangent Approximation based Variational Inference Framework for Strongly Super-Gaussian Likelihoods ​

Author: Somjit Roy, Pritam Dey, Debdeep Pati, Bani K. Mallick
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, math.ST, stat.CO, stat.ML, stat.TH

arXiv:2504.05431v3 Announce Type: cross Abstract: Variational inference, as an alternative to Markov chain Monte Carlo sampling, has played a transformative role in enabling scalable computation for complex Bayesian models. Nevertheless, existing approaches often depend on either rigid model-specifi...

📖 Read original article


78. Probabilistic Symbolic Regression for Equation Discovery via Operator-induced and Regularized Symbolic Forests ​

Author: Somjit Roy, Pritam Dey, Bani K. Mallick, Debdeep Pati
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, cs.SC, math.ST, stat.ML, stat.TH

arXiv:2509.19710v2 Announce Type: cross Abstract: Symbolic regression has emerged as a powerful tool for artificial intelligence-driven scientific discovery by learning interpretable analytical expressions that reveal governing relationships directly from data. Existing methods, however, often rely ...

📖 Read original article


79. Exploring Line Bundle Standard Models with Transformers ​

Author: Jacky H. T. Yip, Alessandro Mininno, Gary Shiu
Published: 7/29/2026, 4:00:00 AM
Categories: hep-th, cs.LG

arXiv:2607.00078v2 Announce Type: cross Abstract: We propose a Transformer-based Reinforcement Learning architecture, "LB-Explorer", to search for heterotic line bundle standard models arising from compactifications on smooth Calabi-Yau (CY) threefolds. We focus on $E_8\times E_8$ heterotic string t...

📖 Read original article


80. RoCo-ACE: Rollout-Conditioned Online Distillation for Retention-Aware Knowledge Injection ​

Author: Yan Hong, Wei Li, Kedong Xiu, Jun Lan, Shuheng Zhou, Zhongcai Lyu, Huijia Zhu, Weiqiang Wang, Jianfu Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.24771v1 Announce Type: cross Abstract: Knowledge injection updates pretrained MLLMs with new factual or domain-specific knowledge, but fitting full authoritative answers can cause drift in non-updated behavior. Online distillation mitigates this drift by training on model-generated rollou...

📖 Read original article


81. JKO-RAG: Distributional Retrieval as Wasserstein Free-Energy Gradient Flow ​

Author: Levi Segal, Murari Ambati
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.24776v1 Announce Type: cross Abstract: RAG pipelines return a \emph{ranked list} of passages. We argue this is a mismatch: the downstream language model conditions on a \emph{set}, and the selection problem is fundamentally geometric. We propose \jko, which frames reranking as minimising ...

📖 Read original article


82. Steering topology distributions for unified generative design of architected metamaterials ​

Author: Haolin Li, Yuyang Miao, Menglei Li, Jinshuai Bai, Liyuan Wang, Xin Liu, Bo Gao, Jiantao Liu, Danilo Mandic, Zahra Sharif Khodaei, M. H. Aliabadi, Weiqiu Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cond-mat.mtrl-sci, cs.LG

arXiv:2607.24777v1 Announce Type: cross Abstract: Architected metamaterials derive their functions from structure, creating vast opportunities to program physical responses through topology design. However, existing design methods are often tailored to individual design problems, making limited use ...

📖 Read original article


83. HOBA: Hierarchical On-Policy Bidding Agents for Adaptive Online Advertising ​

Author: Ji Wu, Yunshan Peng, Wentao Bai, Yunke Bai, Wenzheng Shu, Jinan Pang, Yanxiang Zeng, Xialong Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.24779v1 Announce Type: cross Abstract: Online advertising bidding systems typically deploy multiple offline-trained expert models (e.g., PID controllers, model predictive control, offline RL policies) but face two critical limitations: lack of online adaptability to non-stationary auction...

📖 Read original article


84. SpecPrefetch: Parameter-Efficient Expert Prefetching for Sparse MoE Foundation Models ​

Author: Jinwei Kong, Runqi Meng, Fanyi Wang, Wentao Qiu, Haotian Hu, Yongjian Zhou, Zhenhua Ge
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.24787v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) models expand foundation model capacity through conditional expert activation, but their full expert pools remain difficult to deploy under limited accelerator memory. Although expert offloading alleviates memory press...

📖 Read original article


85. GLIDE: Guided Layerwise Hybrid Attention for Efficient LLM Inference ​

Author: Vimal William, Ravi Tandon, Jyotikrishna Dass
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.24788v1 Announce Type: cross Abstract: As Large Language Models scale to increasingly long contexts, the memory I/O and computational overhead of the Key-Value (KV) cache during decoding emerges as the primary throughput bottleneck. To address this, we propose GLIDE, a Guided Layerwise Hy...

📖 Read original article


86. NEXT: Reasoning-Driven Video Recommendation via a Vision-Language Model ​

Author: Yuming Liu, Hongye Yang, Harrison Zhao, Ellie Zhu, Bokai Cao, Lei Huang, Lizhu Zhang, Xiangjun Fan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.CV, cs.LG, cs.MM

arXiv:2607.24789v1 Announce Type: cross Abstract: We present NEXT (Next-interest EXploration Transformer), a reasoning-driven video recommendation framework that reasons over the video a user has just watched, infers the viewer's next intent, and retrieves concrete follow-up videos. Explicit continu...

📖 Read original article


87. A GAN-Based Framework for Robust Data Synthesis in Satellite Internet Observations ​

Author: Xiang Shi, Peng Hu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NI

arXiv:2607.24790v1 Announce Type: cross Abstract: Low-Earth orbit (LEO) satellite Internet has become an important infrastructure for enabling ubiquitous connectivity to align with the International Telecommunications Union vision for 6G telecommunications networks. However, current LEO satellite In...

📖 Read original article


88. Selective Impairment of Motor Recovery from Typing Errors in Parkinson's Disease: A Survival Analysis ​

Author: Navin Bondade
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.LG

arXiv:2607.24796v1 Announce Type: cross Abstract: Parkinson's disease (PD) affects multiple, dissociable stages of motor and cognitive control. We ask whether passively-collected keystroke dynamics can distinguish two of these stages: noticing a self-generated error (error monitoring) versus recover...

📖 Read original article


89. Reading Without a Reader: Large Language Models Collapse Reading and Writing into a Single Entangled Code ​

Author: Diego Salda~na Ulloa
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.CL, cs.LG

arXiv:2607.24797v1 Announce Type: cross Abstract: In the literate human brain, reading and writing are two doubly-dissociable systems: a ventral decoding route (impaired in pure alexia) and a fronto-parietal encoding route (impaired in pure agraphia), sharing a partial orthographic core. A decoder-o...

📖 Read original article


90. Bumblebee: Interleaved Mixed-Layer Building Blocks for Large-Scale Recommendation Systems ​

Author: David Bauer, Cancan Zhang, Wenshun Liu, Xiaoyi Zhang, Weijia Liu, Wanli Ma, Yue Weng, Wei Li, Rui Li, Jing Qian, Huayu Li, Xiaoyi Liu, Linhong Zhu, Jerry Fu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.24804v1 Announce Type: cross Abstract: Recommendation systems have undergone significant transformations in the past years. The transition from traditional feature interaction modules to generative next-action prediction has pushed the boundaries of personalized content. Developments have...

📖 Read original article


91. Forgetting Is Not a Fix: Path Dependence in Sequential Engram Editing ​

Author: Ferdinand M. Schessl
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG

arXiv:2607.24805v1 Announce Type: cross Abstract: AI Engram (Kwon et al., 2026) formalizes the four engram criteria of neuroscience as a constrained inverse problem in weight space and solves it closed-form: concept-specific memory traces become linear objects that can be extracted once and combined...

📖 Read original article


92. Towards Real-World RUL Prediction for Aircraft Compressors Under Variable Conditions with Physics-Guided Domain Adaptation ​

Author: Yang Zhang, Shashvat Prakash, Jiong Tang
Published: 7/29/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2607.24809v1 Announce Type: cross Abstract: Remaining useful life prediction for aircraft centrifugal air compressors in real commercial operations poses challenges that controlled benchmark datasets do not expose. In-flight sensor signals superimpose genuine degradation on continuously varyin...

📖 Read original article


93. Aletheia: An Offline-First Clinical Decision Support System for Differential Diagnosis in Low-Resource Healthcare Settings ​

Author: Joseph Walusimbi, Ann Move Oguti, Abubakhari Sserwadda, Precious Boss Kasasira, Charles Brian Okoboi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, q-bio.OT

arXiv:2607.24814v1 Announce Type: cross Abstract: Access to specialist clinical expertise remains severely limited across sub-Saharan Africa, where physician-to-patient ratios can fall below 1:25,000 in rural settings. Existing AI-assisted diagnostic tools predominantly require reliable internet con...

📖 Read original article


94. Dynamic Multi-Criteria Bottleneck Severity Index (DMBSI) for Semiconductor Wafer Manufacturing: A Genetically Optimised Framework for Reentrant Production Systems ​

Author: Mohammad Sharifur Rahman, Karl McCreadie, Saugat Bhattacharyya, M M Manjurul Islam, Cormac McAteer, Bryan John Baker, Nuala Parker, Girijesh Prasad
Published: 7/29/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG

arXiv:2607.24819v1 Announce Type: cross Abstract: Wafer fabrication exhibits unique characteristics, including reentrant process flows, variable bottlenecks, and highly variable process conditions. In order to identify the most severe bottleneck at each moment in time for semiconductor wafer fabrica...

📖 Read original article


95. Stronger Memory-Query Tradeoffs for Convex Optimization: The Limitations of Subquadratic Memory ​

Author: Michael Menart, Aleksandar Nikolov, Ohad Shamir
Published: 7/29/2026, 4:00:00 AM
Categories: cs.DS, cs.CC, cs.LG, math.OC

arXiv:2607.24827v1 Announce Type: cross Abstract: We prove two lower bounds for the first order oracle complexity of minimizing a $d$-dimensional $1$-Lipschitz convex function over the unit ball with $m$ bits of memory. We first show that any such (possibly randomized) algorithm must make $\tilde{\O...

📖 Read original article


96. Improving Rare Medication Recommendation with Counterfactual Data Augmentation and Large Language Models ​

Author: Shinhwan Kang, Soo Yong Lee, Jaewon Kim, Kijung Shin, Buru Chang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.24829v1 Announce Type: cross Abstract: AI-based medication recommendation systems have attracted substantial attention due to their potential to enhance patient safety and therapeutic outcomes. Despite the clinical importance of accurately recommending rarely prescribed medications (rare-...

📖 Read original article


97. Foundation Models for EEG Are Blind to Long-Range Temporal Correlations: A Spectral-Temporal Dissociation Behind Their Cross-Population Fragility ​

Author: Marzieh Zare
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.NC, cs.AI, cs.ET, cs.LG

arXiv:2607.24834v1 Announce Type: cross Abstract: Objective. Electroencephalography (EEG) foundation models (FMs) are trained to reconstruct or contrastively align short patches, then pooled into a fixed embedding. We tested whether these embeddings retained the long-range temporal correlations (LRT...

📖 Read original article


98. Gradient-Based Latent Decomposition Reveals Mechanisms of Feature Degradation in Weakly Supervised Mammography ​

Author: Vinceline Bertrand, Ionut Cardei
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.24835v1 Announce Type: cross Abstract: Weakly supervised hierarchical models exhibit a persistent asymmetry: coarse lesion-type features are preserved under reconstruction while fine-grained malignancy cues degrade---a pattern with direct consequences for the clinical reliability of breas...

📖 Read original article


99. Neuromorphic Diffusion Language Models: Addressing Compute and Memory Bottlenecks via Sparsity and Block Denoising ​

Author: Dengyu Wu, Clement Ruah, Jiechen Chen, Bipin Rajendran, Osvaldo Simeone
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, eess.SP

arXiv:2607.24841v1 Announce Type: cross Abstract: Autoregressive (AR) large language models (LLMs) are inherently inefficient at inference time because each generated token requires accessing the full set of model parameters, leading to low operational intensity and high energy consumption. Masked d...

📖 Read original article


100. Beyond Predictive Accuracy: A Reliability-Aware Audit of Molecular Representations for Human Olfaction ​

Author: Kai Lun Huang (California State University, Fullerton), Wei Chieh Sun (University of Washington)
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG

arXiv:2607.24848v1 Announce Type: cross Abstract: Pretrained molecular encoders are commonly evaluated through downstream prediction, but predictive accuracy alone does not establish that a learned representation captures reproducible scientific structure, adds information beyond strong conventional...

📖 Read original article


101. SearchArt: Training Long-Horizon Search Agent with Scalable Synthetic and Verified Task ​

Author: Lang Mei, Xiaohan Yu, Chong Chen, Liyan Liu, Xiangnan Chen, Jinchao Ma, Chao Feng, Li Huang, Siyu Mo, Sichen Kang, Yunkun Xu, Zhihan Yang, Zhujun Xue, Jingren Zhang, Qing He, Yingdi Huang, Hao Jiang, Ziao Ma, Zewei Pan, Minhao Sun, Zhuo Tao, Jinzhao Xiao, Gangtao Xin, Huanyao Zhang, Wenjian Zhang, Jiangshan Zhang, Guojie Zhu, Jiaxin Mao, Wentao Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.24850v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled search agents to autonomously tackle complex tasks across extended search and reasoning horizons. However, training effective search agents remains challenging due to the lack of scalable a...

📖 Read original article


102. Tokens are All You Need: Dual-purpose Semantic IDs for Achieving LLM-Level I/O Efficiency in recommendation systems ​

Author: Baolei Li, Yiping Yuan, Yilin Zheng, Likang Yin, Ling Liu, Fabio Soldo, Romer Rosales, Xinyang Yi, Lichan Hong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2607.24865v1 Announce Type: cross Abstract: Large-scale recommendation systems face "Memory Wall" bottlenecks due to massive, dense embedding tables. While generative retrieval uses discrete tokens for IDs, high-dimensional context still relies on inefficient dense formats. Inspired by compute...

📖 Read original article


103. Ranked by Position: Order Sensitivity as an Exploitable Attack Surface in LLM Listwise Recommenders ​

Author: Ge Zhang, Jingru Cheng, Huiyuan Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.CL, cs.LG

arXiv:2607.24869v1 Announce Type: cross Abstract: Large language models (LLMs) used as listwise rerankers in recommendation systems suffer from position bias when serializing candidate sets into prompts. We show this order sensitivity creates an exploitable attack surface: an attacker can promote a ...

📖 Read original article


104. GraphRareBench: An Auditable Graph-Evidence Benchmark for Phenotype-Driven Rare-Disease Diagnosis ​

Author: Guiling Guo, Jia Yang, Jiahao Xu, Shuyuan Zheng, Zhonghai Sun, Qiyuan Li
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG

arXiv:2607.24878v1 Announce Type: cross Abstract: Phenotype-driven diagnostic benchmarks usually report the rank of the reference disease, but they rarely reveal which plausible alternatives are ranked above it or what evidence a tool-using model examines before making its decision. We introduce Gra...

📖 Read original article


105. Beyond "What to Retrieve": Uncertainty in Retrieval-Augmented Code Generation ​

Author: Chandan Kumar Sah, Li Zhang, Xiaoli Lian
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CL, cs.LG

arXiv:2607.24884v2 Announce Type: cross Abstract: Repository-level code generation relies on heterogeneous evidence whose relevance, compatibility, and completeness are inherently uncertain. Similar-code examples, repository context, and project-specific APIs may provide complementary information, b...

📖 Read original article


106. Amortising Trajectory Optimisation for Residual MPC via Implicit Contact Differentiation ​

Author: Daniel Layeghi, Thomas Corb`{e}res, Calum Arnott, Aditya Kamireddypalli, Hashim Al-Obaidi, Steve Tonneau, Michael Mistry
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.DC, cs.LG, cs.SY, eess.SY, math.OC

arXiv:2607.24959v1 Announce Type: cross Abstract: Differentiable simulation can accelerate contact-rich trajectory optimisation by exposing local sensitivities of task outcomes to controls. Existing approaches either use finite differences, which are expensive and step-size sensitive; differentiate ...

📖 Read original article


107. Fast, accurate, and differentiable: a neural-network surrogate for NRSur7dq4 precessing binary black hole waveforms ​

Author: Michael P"urrer, Ashwin Girish, Lucy M. Thomas, Scott E. Field, Vijay Varma
Published: 7/29/2026, 4:00:00 AM
Categories: gr-qc, astro-ph.HE, astro-ph.IM, cs.LG

arXiv:2607.24960v1 Announce Type: cross Abstract: We present a neural network surrogate model that emulates the NRSur7dq4 gravitational waveform model for precessing binary black hole mergers. The surrogate decomposes the waveform into constituent quantities and trains an independent multilayer perc...

📖 Read original article


108. Automatic Knowledge Graph Construction and Query for Earthquake Catalogs ​

Author: Yuxin Zhou, Huai Zhang, S. Mostafa Mousavi
Published: 7/29/2026, 4:00:00 AM
Categories: physics.geo-ph, cs.AI, cs.LG

arXiv:2607.24984v1 Announce Type: cross Abstract: In recent years, the number of events in earthquake catalogs has significantly increased due to the utilization of more effective deep learning based detectors and phase pickers but answering open ended questions such as what characterizes this seque...

📖 Read original article


109. Simulation-based parameter estimation via a combination of embedded normalizing flows and implied empirical probabilities under moment restrictions ​

Author: Getachew K. Befekadu
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ME, cs.LG

arXiv:2607.25026v1 Announce Type: cross Abstract: In this work, we present a simulation-based parameter estimation framework for a model defined by a computational simulation of a physical system. We specifically outline an estimation framework consisting of two closely-integrated steps that facilit...

📖 Read original article


110. ScoreShield: Differentially Private Release of Similarity Scores ​

Author: Behrooz Razeghi, Parsa Rahimi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.CR, cs.CV, cs.LG

arXiv:2607.25041v1 Announce Type: cross Abstract: A growing number of applications, such as biometrics and retrieval-augmented generation (RAG), rely on cosine similarity scores computed between vector embeddings of text, images, or audio. These systems return similarity scores through their APIs fo...

📖 Read original article


111. Characterizing and Mitigating the Effects of Device Temperature on RF Fingerprinting Accuracy ​

Author: Haytham Albousayri, Bechir Hamdaoui
Published: 7/29/2026, 4:00:00 AM
Categories: eess.SP, cs.CR, cs.LG

arXiv:2607.25070v1 Announce Type: cross Abstract: Radio Frequency Fingerprinting (RFFP) has emerged as a promising approach for device authentication by exploiting hardware-specific impairments embedded in transmitted signals. Yet existing methods largely overlook a major drawback: RFFP sensitivity ...

📖 Read original article


112. Spectral Truncation in Synthetic Control ​

Author: Mojtaba Eslami
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ME, cs.AI, cs.LG, econ.EM, stat.AP

arXiv:2607.25074v1 Announce Type: cross Abstract: Synthetic control (SC) matches a treated unit's pre-treatment trajectory to a weighted combination of donor units. We study Spectral SC, which instead matches the treated unit in coordinates defined by the leading temporal singular vectors of the don...

📖 Read original article


113. Matryoshka Agent: Unfolding Sub-Agents for Long-Horizon Machine Learning Engineering ​

Author: Rushi Qiang, Changhao Li, Haotian Sun, Yuchen Zhuang, Chao Zhang, Bo Dai
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.25090v1 Announce Type: cross Abstract: Machine learning engineering (MLE) tasks require long-horizon decision making over iterative solution debugging and refinement, under expensive and feedback-driven environment interactions. Developing and training a monolithic agent for such tasks is...

📖 Read original article


114. Towards Robust Reinforcement Learning for Small-Scale Language Model Agents ​

Author: Md Rezwanul Haque, Md. Milon Islam, Fakhri Karray
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, math.OC, stat.CO

arXiv:2607.25091v1 Announce Type: cross Abstract: The alignment of Small Language Models (SLMs) in the 70--500M parameter range using reinforcement learning is often considered unstable, though the underlying failure mechanisms have not been systematically investigated. In the State-of-the-Art (SOTA...

📖 Read original article


115. MOSAIC-FL, a micro-service based privacy-preserving framework with application to genomics ​

Author: Paul Largillier, Karl Paygambar, C'edric Gouy-Pailler, Vincent Meyer, Mallek Mziou, Oana Stan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2607.25107v1 Announce Type: cross Abstract: Security and privacy are primordial requirements for Federated Learning (FL), especially in fields such as healthcare and genomics where sensitive information has to be analyzed. Our FL framework is designed to address these challenges while proposin...

📖 Read original article


116. OPERA: Offline Policy-guided Expert Routing and Adaptation for Universal Biomedical Image Analysis ​

Author: Zihan Li, Feiyang Liu, Dandan Shan, Ruibo Wang, Qingqi Hong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV

arXiv:2607.25108v1 Announce Type: cross Abstract: Biomedical image analysis spans diverse modalities and tasks, yet real-world deployment is hindered by severe distribution shifts across scanners, protocols, and patient populations. High-performing models consequently require repeated domain-specifi...

📖 Read original article


117. Memory Layer: Train the In-Model Cache for Recommendation Models ​

Author: Liangyuan Na, Gufan Yin, Yixin Bao, Xianjie Chen, Justin Lin, Ziheng huang, Xinyuan Zhang, Wen Zhang, Hao Lin, Xiaoheng Mao, Shuo Tang, Min Yu, Lei Chen, Chao yang, Ziliang Zhao, Mengjiao Zhou, Zheng Qi, Dmitry Barablin, Chuo-Yun Yang, Kaustubh Vartak, Tingting Zhang, Arun Kumar Singh
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.25110v1 Announce Type: cross Abstract: Early ranking stages in recommendation systems precompute item embeddings and cache them in-model for scoring within strict latency constraints. Because this cache exists only at serving time, outside the training loop, training and serving use diffe...

📖 Read original article


118. Analysis of the Shortcut Learning and Clever Hans Effect in CNN based ECG Image Classification ​

Author: Abhay Kumar Pathak, Mrityunjay Chaubey, Manjari Gupta, Deepti Mishra
Published: 7/29/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG

arXiv:2607.25117v1 Announce Type: cross Abstract: Deep learning models for ECG image classification may achieve high accuracy by exploiting non-physiological visual cues instead of ECG waveform morphology. Given the black-box nature of deep learning models, their promise of high predictive performan...

📖 Read original article


119. Deep Label-Wise Attentive Temporal Convolutional Networks Improve Medical Coding ​

Author: Muhammed Yavuz Nuzumlal{\i}, Alexander Fabbri, Irene Li, Dragomir Radev
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.25129v1 Announce Type: cross Abstract: Medical coding is the task of assigning a set of diagnosis and procedure codes for a hospitalization using recorded notes. It requires aggregating information from different parts of the text and focus to different sections for each individual code, ...

📖 Read original article


120. Learning from 53.6K Real-World Developer Edits of AI-Generated Code ​

Author: Jenny T. Liang, Mihika Bairathi, Wayne Chi, Ameet Talwalkar, Nishant Subramani, Valerie Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.HC, cs.LG

arXiv:2607.25130v1 Announce Type: cross Abstract: Imperfections in AI-generated code require that software developers modify the generated code manually, or by re-prompting an AI programming assistant. Manual code edits provide more realistic and granular information on editing behavior than Git com...

📖 Read original article


121. ScalableRAG: High-Quality RAG at Zero Ingestion Cost ​

Author: Hilaf Hasson, Aditya Chakravarty, Jayant Thomas, Krishna Gogineni
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.25135v1 Announce Type: cross Abstract: Recent advances in RAG aim to optimize for performance by paying high ingestion costs for knowledge ingestion: building knowledge graphs or extracting SQL tables. In this work we show that the operations that such knowledge bases allow can be replica...

📖 Read original article


122. A Foundational Perspective for Partitional Clustering on Networks ​

Author: Derya Ipek Eroglu, Cem Iyigun
Published: 7/29/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2607.25144v1 Announce Type: cross Abstract: This study presents a theoretical analysis of partitional clustering on networks, analyzing both hard and soft assignment schemes with different objective functions. Cluster centers are not restricted to vertices but can also be located along the edg...

📖 Read original article


123. A Riemannian View on Active Subspaces ​

Author: Zachary Grey
Published: 7/29/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.DG

arXiv:2607.25163v1 Announce Type: cross Abstract: Active subspaces provide an explainable, eigenvalue-ordered principle for studying how scalar-valued quantities of interest change the most, on average, over a reduced basis of Euclidean domains. Composition with parallel transport generalizes this p...

📖 Read original article


124. Lloyd's $K$-Means Clustering Algorithm Is Frank-Wolfe in Disguise ​

Author: Michael Pokojovy, J. Marcus Jobe, Simon Lacoste-Julien
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO

arXiv:2607.25190v1 Announce Type: cross Abstract: Lloyd's $K$-means algorithm, also known as na"{i}ve $K$-means, is a widely used ad hoc optimization heuristic, designed to minimize the sum of squared errors (SSE) across all $K$-partitions of a dataset via iterative cluster refinement. In this work...

📖 Read original article


125. SecDrift: Measuring Sector-Conditioned Security Drift in AI-Generated Code ​

Author: Narayanaswami Natraj Bharadwaj, Dhivya Chandramouleeswaran
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.LG, cs.SE

arXiv:2607.25225v1 Announce Type: cross Abstract: LLMs are increasingly used for code generation in critical infrastructure, yet the security effect of domain-specific prompting is understudied. We present SecDrift, a benchmark measuring sector-conditioned security drift: the change in static-analys...

📖 Read original article


126. Decision-Level Hijacking: Injecting Cognitive Bias into Large Language Models via Bit-Flip Attacks ​

Author: Yu Yan, Jiahao Chen, Siqi Lu, Yongjuan Wang, Ziming Zhao, Zhaoxuan Li, Tianyu Du, Qingjun Yuan, Shouling Ji
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2607.25227v1 Announce Type: cross Abstract: Large Language Models (LLMs) have been widely applied in high-stakes decision-making scenarios such as corporate strategy, and users are increasingly relying on their outputs. However, the deep integration of open-source model sharing ecosystems with...

📖 Read original article


127. Learning from the Unseen: Offline Reinforcement Learning with Hidden Actions ​

Author: Zeyu Bian, Ying Zhou, Yifan Cui
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.25241v1 Announce Type: cross Abstract: Standard offline reinforcement learning (RL) algorithms typically assume that the actions in the dataset are observed without error. However, in many real-world applications, the true actions are unobserved and only noisy proxies are available, causi...

📖 Read original article


128. Laplace-PSN-IRT: Uncertainty Quantification for Neural Item Response Theory Models of LLM Benchmarks ​

Author: Juan Francisco, Mandujano Reyes
Published: 7/29/2026, 4:00:00 AM
Categories: stat.AP, cs.AI, cs.LG

arXiv:2607.25257v1 Announce Type: cross Abstract: Item Response Theory (IRT) has recently been proposed as a framework for evaluating large language model (LLM) benchmarks by separating a model's latent ability from the properties of individual benchmark items. Existing neural IRT approaches, includ...

📖 Read original article


129. FORGE: Frame Orthogonality in Relevance Geometry for Long-Form Video Understanding ​

Author: Ghazal Kaviani, Ghassan AlRegib
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.25266v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have enabled long-form video understanding at a scale that was not previously possible. However, the density of relevant content decreases sharply as video sequence length increases, and exposing the model to ...

📖 Read original article


130. Where Steering Signals Come From: Activation Source Selection in Activation Steering ​

Author: Jiaran Ye, Lingxu Ran, Zijun Yao, Chenpeng Wang, Yong Jiang, Lei Hou, Juanzi Li, Liangming Pan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.25270v1 Announce Type: cross Abstract: Activation steering controls language models by adding vectors or features to hidden states at inference time, but the upstream source of these steering signals is often treated as a secondary detail. We study this source choice as activation source ...

📖 Read original article


131. FunnelAL: Retrieve-then-Rank Active Learning for Single-Class Discovery ​

Author: Reihaneh Rostami (RAIC Labs), Brian Goodwin (RAIC Labs)
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.IR, cs.LG

arXiv:2607.25276v1 Announce Type: cross Abstract: We present FunnelAL, a retrieve-then-rank active learning system for single-class discovery, which adapts the multi-stage funnel architecture of industrial recommender systems to data annotation. Large-scale supervised learning faces two challenges: ...

📖 Read original article


132. Normalizing Flows to Reconstruct Pseudo-PDFs ​

Author: Yamil Cahuana Medrano, Kostas Orginos
Published: 7/29/2026, 4:00:00 AM
Categories: hep-lat, cs.LG, hep-ph

arXiv:2607.25282v1 Announce Type: cross Abstract: We investigate a normalizing-flow approach for reconstructing parton distribution functions (PDFs) from synthetic matrix-element data. Our framework combines Gaussian Process priors with invertible neural networks to learn a posterior distribution ov...

📖 Read original article


133. CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition ​

Author: Lai Wei, Chengqi Li, Jiapeng Li, Ruina Hu, Yue Wang, Weiran Huang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.CL, cs.LG

arXiv:2607.25294v1 Announce Type: cross Abstract: Real-world tasks often require models to learn from task-specific context rather than relying only on pre-trained knowledge. While recent work has highlighted this capability as context learning, existing evaluations mainly focus on textual contexts....

📖 Read original article


134. Physics-Informed Neural Operator for Warm-Starting Background-Decomposed and Preconditioned PSFD: Enabling Scalable 3-D EUV Mask Simulation ​

Author: Doyun Kim, Werner Gillijns
Published: 7/29/2026, 4:00:00 AM
Categories: physics.optics, cs.AI, cs.LG, physics.app-ph

arXiv:2607.25330v1 Announce Type: cross Abstract: We present a physics-informed neural operator (PINO) trained with pseudo-spectral frequency-domain (PSFD) equations for electromagnetic (EM) scattering problems in EUV lithography. The Fourier neural operator is factorized into a two-dimensional late...

📖 Read original article


135. Sharpness-aware Model Merging with Salience Recovery for LLM-based Cross-Domain Sequential Recommendation ​

Author: Huwei Ji, Jiajie Su, Yuyuan Li, Xiaohua Feng, Chaochao Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.25366v1 Announce Type: cross Abstract: LLM-based Cross-Domain Sequential Recommendation (CDSR) leverages LLMs to enhance target performance via deep semantic reasoning, alleviating the dependency on overlapping users. Among LLM-based paradigms, model merging is particularly promising for ...

📖 Read original article


136. Robust Unsupervised Network Intrusion Detection via Federated Learning with Selective Aggregation under Anomalous Sample Contamination ​

Author: Shohei Kamiguchi, Takayuki Nishio
Published: 7/29/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2607.25439v1 Announce Type: cross Abstract: Network intrusion detection systems (NIDS) have become essential for Internet of Things (IoT) environments, as malware targeting IoT devices continues to evolve in sophistication. Unsupervised learning approaches offer a promising direction by removi...

📖 Read original article


137. Toward an Organizational Science of Multi-Agent LLM Systems: Decoupling Who, How, and Which Algorithm ​

Author: Huan Chen, Xiang Song, Jian Jin, Pan Ren, Liang-Jie Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.25446v1 Announce Type: cross Abstract: Multi-agent frameworks built on large language models (LLMs) routinely entangle three logically distinct concerns: who is on the team (organization), how members align (coordination), and which algorithm fuses their work (collaboration protocol). IMA...

📖 Read original article


138. Architectural Backdoors in Vision-Language Model Supply Chains via Representation Steering ​

Author: Maria Rosaria Briglia, Igor Maljkovic, Antonio Emanuele Cin`a, Luca Oneto, Iacopo Masi, Fabio Roli
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2607.25479v1 Announce Type: cross Abstract: Vision--Language Models (VLMs) are increasingly deployed through a model supply chain in which pretrained checkpoints, architecture definitions, text encoders, and exported computation graphs are distributed by third parties and reused across downstr...

📖 Read original article


139. Shared Voxel-Map-Based Cooperative Indoor UAV Guidance with a Multi-Agent Soft Actor-Critic Controller ​

Author: Thomas Hickling, Dylan Wynne, Yu Su, Nabil Aouf
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.25728v1 Announce Type: cross Abstract: This paper presents a cooperative indoor UAV guidance framework that combines a shared voxel-map world model with a multi-agent Soft Actor-Critic (MASAC) controller. Multiple drones fuse 360 LiDAR observations into a common world-frame occupancy map,...

📖 Read original article


140. WALoMA: A Multitask Wireless Foundation Model via Adaptive Low-Rank Masked Autoencoders ​

Author: Madi Makin, Asmaa Abdallah, Abdulkadir Celik, Ahmed M. Eltawil
Published: 7/29/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2607.25763v1 Announce Type: cross Abstract: This paper proposes a multitask wireless foundation model via adaptive low-rank masked autoencoders (WALoMA), a unified multi-task foundation model for sixth-generation (6G) wireless physical layer architectures, to address the limitations of special...

📖 Read original article


141. Variance-Reduced Conditional Gradient Methods under Markovian Sampling for Nonconvex Composite Optimization ​

Author: Zhaojun Peng
Published: 7/29/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2607.25785v1 Announce Type: cross Abstract: We study stochastic composite nonconvex optimization over a compact convex set when gradient samples arrive along a single trajectory of a fixed ergodic Markov chain. Existing single-trajectory variance-reduction theory covers smooth unconstrained ob...

📖 Read original article


142. VAD to the Bone: Ultra-Tiny Speech Activity Detection for Edge Deployment ​

Author: Stephen Bauer, Sheila Seidel, Shanza Iftikhar, Scott Veidenheimer, Gorkem Ulkar
Published: 7/29/2026, 4:00:00 AM
Categories: eess.AS, cs.LG

arXiv:2607.25870v1 Announce Type: cross Abstract: Voice activity detection (VAD) triggers downstream speech processing in always-on systems under strict memory, latency, and compute constraints. Recent compact models report strong accuracy but rely on components that are not widely supported: learna...

📖 Read original article


143. HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone ​

Author: Simple AI, :, Yuteng Wei, Jinming Ma, Jiawei Wang, Weitao Zhou, Yushen Zuo, Ke Rui, Minglei Li, Jinhao Zhang, Zhikang Pan, Xiang Wang, Haoran Jia, Huan Du, Zicheng Zeng, Jun Ma, Guiyu Qin, Di Zhang, Xiaofei Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2607.25895v1 Announce Type: cross Abstract: Learning deployable manipulation policies is bottlenecked by the scarcity of data that is both high-fidelity and scalable. Real-robot teleoperation is accurate but costly to scale; robot-free UMI capture scales readily, and current practice uses the ...

📖 Read original article


144. Can Deep Generative Models Reproduce Non-Stationary Gaussian Random Fields? ​

Author: Daniel Kua, Yan Song
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.25929v1 Announce Type: cross Abstract: Deep generative models (DGMs) are widely used for complex high-dimensional data and increasingly applied to spatial and spatio-temporal modeling. Their generated samples implicitly represent the learned data distribution and associated uncertainty. H...

📖 Read original article


145. MODUS: Decoder-Only Any-to-Any Modeling of Diverse Modalities ​

Author: Mingqiao Ye, Zhaochong An, Zhitong Gao, Xian Liu, Fran\c{c}ois Fleuret, Chuan Li, Amir Zadeh, Serge Belongie, Afshin Dehghan, Jesse Allardice, David Mizrahi, O\u{g}uzhan Fatih Kar, Roman Bachmann, Amir Zamir
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.25948v1 Announce Type: cross Abstract: Any-to-any models predict any modality from any combination of others within a single network, a formulation used in multimodal vision and vision-language models, and increasingly in scientific domains such as ecology and astronomy. Existing any-to-a...

📖 Read original article


146. Quasi-SVD: Learning a Lie-constrained matrix factorisation for real-time imaging ​

Author: Christopher Hahne
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.NA, math.NA

arXiv:2607.25967v1 Announce Type: cross Abstract: Singular Value Decomposition (SVD) underlies matrix factorisation tasks across computational imaging, with medical applications increasingly demanding real-time processing. Yet SVD algorithms are inherently sequential, constraining real-time GPU thro...

📖 Read original article


147. Schr\"odinger's Cat: Probabilistic Representation and Prediction of Potential Scene Kinematics ​

Author: Timy Phan, Jannik Wiese, Bj"orn Ommer
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.25984v1 Announce Type: cross Abstract: Predicting how a scene may evolve from partial observations requires reasoning about multiple possible futures rather than committing to a single trajectory. Existing approaches either generate appearance-dominated video predictions or sample a small...

📖 Read original article


148. Physics-Aware End-to-End Deep Reinforcement Learning for Quadcopter Control with Actuator Dynamics ​

Author: Ya-Chia Shen, Woei-Leong Chan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.SY, eess.SY

arXiv:2607.25985v1 Announce Type: cross Abstract: Unmanned aerial vehicles (UAVs), particularly quadcopters, present unique challenges for autonomous control due to their underactuated dynamics: only four available control inputs must govern six degrees of freedom. This paper investigates a physics-...

📖 Read original article


149. Untangling Co-Drift: Proactive Multi-Intent Failure Prediction and Root-Cause Disambiguation for Self-Driving Networks ​

Author: Md. Kamrul Hossain, Walid Aljoby
Published: 7/29/2026, 4:00:00 AM
Categories: cs.NI, cs.LG, cs.RO

arXiv:2607.25989v1 Announce Type: cross Abstract: The vision of self-driving networks that monitor, reason, and act upon themselves with minimal human intervention relies on tightly coupled monitoring, analytics, and actuation functions. In this work, we treat these functions as three operational ma...

📖 Read original article


150. Parallel Decoding Distillation for Fast Image and Video Generation ​

Author: Neta Shaul, Chao Liu, Arash Vahdat, Julius Berner
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.26004v1 Announce Type: cross Abstract: Generation in video diffusion or flow models is computationally expensive due to the slow and iterative sampling process. Current state-of-the-art (SOTA) acceleration methods heavily rely on variational score distillation (VSD) and adversarial losses...

📖 Read original article


151. VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening ​

Author: Syed Mhamudul Hasan, Anas AlSobeh, Hussein Zangoti, Abdur R. Shahid
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.26042v1 Announce Type: cross Abstract: We present VetClaw, an edge-cloud multimodal agentic system for early veterinary disease screening. VetClaw uses a camera module as an edge sensing device and sends captured images, together with optional symptom descriptions, to a server-hosted visi...

📖 Read original article


152. $\pi\mathbf{R}^2$: Reactive Real-time Flow Policies ​

Author: Sungjae Park, Shubham Tulsiani
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.26055v1 Announce Type: cross Abstract: Generalist manipulation policies increasingly take the form of action-chunking flow policies built on large pretrained backbones. Such chunks run open-loop, so the policy cannot react to sensory input arriving mid-execution, sacrificing \emph{reactiv...

📖 Read original article


153. Towards a Statistical Understanding of Neural Networks: Beyond the Neural Tangent Kernel Theories ​

Author: Yicheng Li, Haobo Zhang, Jianfa Lai, Qian Lin, Jun S. Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.TH

arXiv:2412.18756v2 Announce Type: replace Abstract: A primary advantage of neural networks lies in their feature learning characteristics, which is challenging to theoretically analyze due to the complexity of their training dynamics. We examine feature learning and its potential benefits for genera...

📖 Read original article


154. COMPOL: A Unified Neural Operator Framework for Scalable Multi-Physics Simulations ​

Author: Junqi Qu, Tao Wang, Yushun Dong, Hewei Tang, Shibo Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2501.17296v4 Announce Type: replace Abstract: Multiphysics simulations play an essential role in accurately modeling complex interactions across diverse scientific and engineering domains Although neural operators especially the Fourier Neural Operator FNO have significantly improved computati...

📖 Read original article


155. Invariance Pair Guidance: Robustness to Spurious Correlations via Corrective Gradients ​

Author: Martin Surner, Abdelmajid Khelil, Ludwig Bothmann
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2502.18975v3 Announce Type: replace Abstract: Machine learning models are inherently bound to the distribution of the training data, often exploiting non-causal shortcuts. As a result, achieving robustness to spurious correlations remains a challenge. While existing approaches rely on data man...

📖 Read original article


156. Cohort-attention Evaluation Metrics for Tied Data ​

Author: Dongjing Jiang, Qingchong Jiao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, stat.ML

arXiv:2503.12755v4 Announce Type: replace Abstract: Artificial intelligence (AI) has significantly improved medical screening accuracy, particularly in cancer detection and risk assessment. However, traditional classification metrics often fail to account for imbalanced data, varying performance acr...

📖 Read original article


157. Diffusion Disambiguation Models for Partial Label Learning ​

Author: Jinfu Fan, Xiaohui Zhong, Kangrui Ren, Jiangnan Li, Linqing Huang, Min Gan, C. L. Philip Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2507.00411v2 Announce Type: replace Abstract: Learning from ambiguous labels is a long-standing problem in practical machine learning applications. The purpose of \emph{partial label learning} (PLL) is to identify the ground-truth label from a set of candidate labels associated with a given in...

📖 Read original article


158. Annotation-Assisted Learning of Treatment Policies From Multimodal Electronic Health Records ​

Author: Henri Arno, Thomas Demeester
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2507.20993v4 Announce Type: replace Abstract: We study how to learn treatment policies from multimodal electronic health records (EHRs) that consist of tabular data and clinical text. These policies can help physicians make better treatment decisions and allocate healthcare resources more effi...

📖 Read original article


159. Towards real-time surrogate-free Bayesian inversion for neutron reflectometry ​

Author: Max D. Champneys, Andrew J. Parnell, Philipp Gutfreund, Maximilian W. A. Skoda, Patrick A. Fairclough, Timothy J. Rogers, Stephanie L. Burg
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci

arXiv:2509.06924v3 Announce Type: replace Abstract: Neutron reflectometry (NR) is a key enabling technology for many areas of scientific development. Although the forward reflectivity model is well-known, inferring the physical properties of a sample from NR data requires the solution of an inverse ...

📖 Read original article


160. CIFNet: An Analytic Neural Learning Framework for Efficient and Calibrated Class-Incremental Learning ​

Author: Alejandro Dopico-Castro, Oscar Fontenla-Romero, Bertha Guijarro-Berdi~nas, Amparo Alonso-Betanzos
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2509.11285v2 Announce Type: replace Abstract: Class-Incremental Learning (CIL) in deep neural networks is conventionally framed as an iterative gradient-based optimization problem, incurring high computational cost, hyperparameter sensitivity, and risk of catastrophic forgetting. In this work,...

📖 Read original article


161. Long-Term PM2.5 Forecasting Using a DTW-Enhanced CNN-GRU Model ​

Author: Amirali Ataee Naeini, Arshia Ataee Naeini, Fatemeh Karami Mohammadi, Omid Ghaffarpasand
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.22863v2 Announce Type: replace Abstract: Reliable long-term forecasting of PM2.5 concentrations is critical for public health early-warning systems, yet existing deep learning approaches struggle to maintain prediction stability beyond 48 hours, especially in cities with sparse monitoring...

📖 Read original article


162. Dataset Poisoning Attacks on Behavioral Cloning Policies ​

Author: Akansha Kalra, Soumil Datta, Ethan Gilmore, Duc La, Guanhong Tao, Daniel S. Brown
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.RO

arXiv:2511.20992v2 Announce Type: replace Abstract: Behavior Cloning (BC) is a popular framework for training sequential decision policies from expert demonstrations via supervised learning. As these policies are increasingly being deployed in the real world, their robustness and potential vulnerabi...

📖 Read original article


163. Robustness Certificates for Neural Networks Against Data Poisoning and Evasion Attacks ​

Author: Sara Taheri, Mahalakshmi Sabanayagam, Debarghya Ghoshdastidar, Majid Zamani
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2512.20865v3 Announce Type: replace Abstract: The increasing use of machine learning in safety-critical domains amplifies the risk of adversarial threats, especially data poisoning attacks that corrupt training data to degrade performance or induce unsafe behavior. Most existing defenses lack ...

📖 Read original article


164. Deep Delta Learning ​

Author: Yifan Zhang, Yifeng Liu, Mengdi Wang, Quanquan Gu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CV

arXiv:2601.00417v4 Announce Type: replace Abstract: Transformer residual streams evolve through additive updates. Although a sufficiently expressive residual block can represent content replacement, standard architectures do not parameterize reading, comparison, and replacement as an explicit residu...

📖 Read original article


165. Benchmarking Deep Learning Models for Raman Spectroscopy Across Open-Source Datasets ​

Author: Adithya Sineesh, Akshita Ramya Kamsali
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.16107v2 Announce Type: replace Abstract: Deep learning classifiers for Raman spectroscopy are increasingly reported to outperform classical chemometric approaches. However, their evaluations are often conducted in isolation or compared against traditional machine learning methods or trivi...

📖 Read original article


166. Vector-Valued Distributional Reinforcement Learning Policy Evaluation: A Hilbert Space Embedding Approach ​

Author: Mehrdad Mohammadi, Qi Zheng, Ruoqing Zhu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2601.18952v2 Announce Type: replace Abstract: We propose an (offline) multi-dimensional distributional reinforcement learning framework (KE-DRL) that leverages Hilbert space mappings to estimate the kernel mean embedding of the multi-dimensional value distribution under a proposed target polic...

📖 Read original article


167. Regime-Adaptive Bayesian Optimization via Dirichlet Process Mixtures of Gaussian Processes ​

Author: Yan Zhang, Xuefeng Liu, Sipeng Chen, Sascha Ranftl, Chong Liu, Shibo Li
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2601.20043v2 Announce Type: replace Abstract: Standard Bayesian Optimization (BO) assumes uniform smoothness across the search space an assumption violated in multi-regime problems such as molecular conformation search through distinct energy basins or drug discovery across heterogeneous molec...

📖 Read original article


168. Linear-LLM-SCM: Benchmarking LLMs for Coefficient Elicitation in Linear-Gaussian Causal Models ​

Author: Kanta Yamaoka, Sumantrak Mukherjee, Thomas G"artner, David Antony Selby, Stefan Konigorski, Eyke H"ullermeier, Viktor Bengs, Sebastian Josef Vollmer
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.10282v2 Announce Type: replace Abstract: Large language models (LLMs) have shown potential in identifying qualitative causal relations, but their ability to perform quantitative causal reasoning---estimating effect sizes that parametrize functional relationships---remains underexplored in...

📖 Read original article


169. Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates ​

Author: Yusen Huo, Changping Wang, Yangru Huang, Jun Zhang, Jie Jiang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.10430v2 Announce Type: replace Abstract: Off-policy policy optimization reuses historical behavior, including negative-advantage samples that suppress known failures. We show that repeated reuse can turn this useful signal into excessive repulsion: as the learner moves away from a histori...

📖 Read original article


170. AdvSynGNN: Structure-Adaptive Graph Neural Nets via Adversarial Synthesis and Self-Corrective Propagation ​

Author: Rong Fu, Chunlei Meng, Shuo Yin, Kun Liu, Simon Fong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.17071v4 Announce Type: replace Abstract: Graph neural networks frequently encounter significant performance degradation when confronted with structural noise or non-homophilous topologies. To address these systemic vulnerabilities, we present AdvSynGNN, a comprehensive architecture design...

📖 Read original article


171. MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs ​

Author: Ziqiao Shang, Ling-Yue Ge, Zian Xu, Zi-Jian Cheng, Shi-Yu Tian, Zhenyu Huang, Wenbo Fu, Weiming Wu, Yang Chen, Xiangwen Zhang, Yulan Hu, Bin Liu, Lan-Zhe Guo
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.18600v5 Announce Type: replace Abstract: Systematically evaluating Multimodal Large Language Models (MLLMs) is essential for advancing Artificial General Intelligence (AGI). Yet existing benchmarks remain inadequate for rigorously measuring their reasoning capabilities under multi-criteri...

📖 Read original article


172. Real-Time Driver Safety Scoring Through Inverse Crash Probability Modeling ​

Author: Joyjit Roy, Samaresh Kumar Singh, Sushanta Das
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC, cs.ET

arXiv:2603.14841v3 Announce Type: replace Abstract: Road crashes remain a leading cause of preventable fatalities. Existing prediction models predominantly produce binary outcomes, which offer limited actionable insights for real-time driver feedback. These approaches often lack continuous risk quan...

📖 Read original article


173. Operational evaluation of data-driven forest fire forecasting models ​

Author: Shahbaz Alvi, Italo Epicoco, Jose Maria Costa Saura
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.25469v2 Announce Type: replace Abstract: A growing body of literature has focused on predicting wildfire occurrence using machine learning methods, capitalizing on high-resolution data and fire predictors that canonical process-based frameworks largely ignore. Standard evaluation metrics ...

📖 Read original article


174. PEANUT: Perturbations by Eigenvector Alignment for Attacking Graph Neural Networks Under Topology-Driven Message Passing ​

Author: Bhavya Kohli, Biplab Sikdar
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.26136v3 Announce Type: replace Abstract: Message Passing Neural Networks (MPNNs) have achieved strong performance on tasks involving relational data. However, small perturbations to graph structure can significantly alter their outputs, raising concerns about their robustness in real-worl...

📖 Read original article


175. Forecasting Ionospheric Irregularities on GNSS Lines of Sight Using Dynamic Graphs with Ephemeris Conditioning ​

Author: Mert Can Turkmen, Eng Leong Tan, Yee Hui Lee
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, physics.geo-ph, physics.space-ph

arXiv:2604.18379v2 Announce Type: replace Abstract: Most data-driven ionospheric models operate on gridded products, which do not preserve the time-varying sampling structure of satellite-based sensing. We instead model the ionosphere as a dynamic graph over ionospheric pierce points, with connectiv...

📖 Read original article


176. Structured Scaling of AI Discovery Across Diverse Scientific Domains ​

Author: Haotian Ye, Haowei Lin, Jingyi Tang, Yizhen Luo, Rahul Thapa, Caiyin Yang, Chang Su, Rui Yang, Ruihua Liu, Rundao Li, Zeyu Li, Pengwei Sun, Chong Gao, Dachao Ding, Guangrong He, Miaolei Zhang, Lina Sun, Wenyang Wang, Yuchen Zhong, Zhuohao Shen, Puheng Li, Pan Lu, Bianxiao Cui, Di He, Jianzhu Ma, Junfeng Li, Hexi Baoyin, Yejin Choi, Stefano Ermon, Xiaowen Chu, Tongyang Li, Yuzhi Xu, James Zou
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.19341v2 Announce Type: replace Abstract: Scientific discovery often requires many cycles of proposing, testing, and refining candidate solutions. Language models can increasingly participate in these loops, but simply generating more attempts does not ensure progress: parallel searches ma...

📖 Read original article


177. Demographic-Aware Transfer Learning for Sleep Stage Classification in Clinical Polysomnography ​

Author: S M Asif Hossain, Shruti Kshirsagar
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.02245v2 Announce Type: replace Abstract: Automated sleep stage classification typically employs a single population-agnostic model, disregarding established demographic variations in sleep architecture. Sleep patterns, however, differ substantially across gender, age, and obstructive slee...

📖 Read original article


178. Convolutional Neural Networks in Vis-NIR Chemometrics: From Contradiction to Conditional Design ​

Author: D'ario Passos
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, physics.optics

arXiv:2605.02636v2 Announce Type: replace Abstract: Near-infrared (NIR and Vis-NIR) spectroscopy is widely used for rapid, non-destructive analysis in food, agriculture, pharmaceuticals, process analytical technology, and bioprocess monitoring. Nevertheless, deep-learning studies in NIR chemometrics...

📖 Read original article


179. Fully Automatic Trace Gas Plume Detection ​

Author: V'it R\r{u}\v{z}i\v{c}ka, David R. Thompson, Jay E. Fahlen, Amanda M. Lopez, Steven Lu, Chuchu Xiang, Holly Bender, Daniel Jensen, Philip G. Brodrick, Jake Lee, Brian Bue, Daniel H. Cusworth, Luis Guanter, Adam Chlus, Andrew Thorpe, Robert O. Green
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.03372v2 Announce Type: replace Abstract: Future imaging spectrometers will expand contemporary data volumes by orders of magnitude, requiring automated methods to upscale labor-intensive detection of trace gas point sources. Here we present a fully-automated approach that achieves operati...

📖 Read original article


180. Complex-Valued Phase-Coherent Transformer ​

Author: Leona Hioki
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.10123v2 Announce Type: replace Abstract: Complex-valued Transformers have largely inherited softmax attention from real-valued architectures. However, row-normalised token competition is not necessarily aligned with phase-preserving computation. In this paper, we introduce the Phase-Coher...

📖 Read original article


181. Efficient Online Conformal Selection with Limited Feedback ​

Author: Sreenivas Gollapudi, Kostas Kollias, Kamesh Munagala, Ali Sinop
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.14953v3 Announce Type: replace Abstract: We address the problem of conformal selection, where an agent must select a low-cost subset of options to ensure that at least one "success" is identified at a pre-specified target rate $\phi$. While traditional online conformal prediction focuses ...

📖 Read original article


182. A Unified Framework for Uncertainty-Aware Explainable Artificial Intelligence: A Case Study in Power Quality Disturbance Classification ​

Author: Yinsong Chen, Samson S. Yu, Zhong Li, Chee Peng Lim
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.21114v2 Announce Type: replace Abstract: Post-hoc explainable AI (XAI) methods usually return one attribution map, even when the model represents uncertainty in its parameters. We define the \emph{explanation distribution} as the distribution of attribution maps obtained from sampled mode...

📖 Read original article


183. Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability ​

Author: Taewoon Kim, Vincent Fran\c{c}ois-Lavet, Michael Cochez
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.22142v4 Announce Type: replace Abstract: Reinforcement learning under partial observability requires deciding what information to retain, yet most memory-based approaches do not explicitly model short-term-to-long-term transfer of symbolic observations. We study this transfer process in a...

📖 Read original article


184. Adaptive Measurement Allocation for Learning Kernelized SVMs Under Noisy Observations ​

Author: Artur Miroszewski
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.22275v2 Announce Type: replace Abstract: Kernel methods are typically formulated under the assumption of exact, noise-free access to the Gram matrix. However, in emerging settings each kernel entry must be inferred from noisy observations, and its accuracy depends on how a limited measure...

📖 Read original article


185. GoQuant: Geometric Orthogonal Residual Projection for Multiplier-Free Power-of-Two Transformer Quantization ​

Author: Maoyang Xiang, Tao Luo, Bo Wang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.26092v5 Announce Type: replace Abstract: The deployment of Large Language Models (LLMs) and Vision Transformers (ViTs) on edge devices is significantly constrained by memory capacity and the critical timing bottlenecks introduced by dense Multiply--Accumulate (MAC) arrays. In the ultra-lo...

📖 Read original article


186. Neural Networks Provably Learn Spectral Representations for Group Composition ​

Author: Jianliang He, Leda Wang, Fengzhuo Zhang, Siyu Chen, Zhuoran Yang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, math.OC, math.RT, math.ST, stat.ML, stat.TH

arXiv:2606.02993v2 Announce Type: replace Abstract: Understanding how structured internal structure emerges during neural network training is central to the study of deep learning. We investigate this phenomenon through the group composition task, where a two-layer neural network is trained to predi...

📖 Read original article


187. Scalable Perturbation Learning for Online Self-Supervised Learning in Echo State Networks ​

Author: Taiki Yamada, Kantaro Fujiwara
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2607.06079v2 Announce Type: replace Abstract: Intelligent systems should not only solve tasks but also adapt under real-world constraints. Autonomous adaptation via self-supervised learning, sequential adaptation via online learning, and memory-efficient implementation via perturbation-based l...

📖 Read original article


188. Prior-matched evaluation of operational Earth-observation classifiers: a three-number reporting method demonstrated on Sentinel-1 internal-wave detection ​

Author: Jo~ao Pinelo, Jo~ao Gon\c{c}alves, Arun Shukla, Adriana Santos-Ferreira
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, eess.IV

arXiv:2607.07146v2 Announce Type: replace Abstract: The Internal Waves Service screens the Sentinel-1 Wave-mode archive for internal solitary waves, routing detections to experts whose adjudication time is the resource the effort exists to conserve. Because attention is the cost of error, precision ...

📖 Read original article


189. Where to Intervene? Benchmarking Fairness-Aware Learning on Differentially Private Synthetic Tabular Data ​

Author: Vin'icius Gabriel Angelozzi, H'eber H. Arcolezi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2607.07471v2 Announce Type: replace Abstract: Machine learning models are increasingly deployed in high-stakes domains, raising concerns about both privacy and fairness. Differential Privacy (DP) has become a gold standard for privacy-preserving data analysis, while fairness-aware mechanisms a...

📖 Read original article


190. EdgeRefine: Privacy-Utility Balance for Graphs via Jaccard Sampling under Edge Differential Privacy ​

Author: Wenxiu Ding, Muzhi Liu, Zheng Yan, Mingjun Wang, Yifan Zhao, Qiao Liu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.08659v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have shown considerable success in learning from graph-structured data, but their use in privacy-sensitive areas remains difficult because graph structure can leak sensitive link information. To satisfy edge-level diffe...

📖 Read original article


191. Reference Traces for Auditing Invisible Weight Updates and Guiding Exact-Budget Protection ​

Author: Zekai Shang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09800v3 Announce Type: replace Abstract: Direct low-precision write-back can erase optimizer proposals, while aggregate update visibility need not identify parameters worth protecting. We study two uses of high-precision reference traces: candidate-matched auditing before a low-precision ...

📖 Read original article


192. RDQ: Residual Distribution Quantization for Large Language Models ​

Author: Prateek Singh
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.10137v2 Announce Type: replace Abstract: Post-training quantization (PTQ) of large language models degrades sharply below 4-bit precision. We identify the root cause as residual stream distributional drift: quantization noise injected at each transformer layer accumulates in the shared re...

📖 Read original article


193. Learning from Local Walks on Dynamic Graphs with Bandit Feedback ​

Author: Sourav Chakraborty, Amit Kiran Rege, Claire Monteleoni, Lijun Chen
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.10571v2 Announce Type: replace Abstract: We study stochastic multi-armed bandits on dynamic graphs, where arms correspond to the vertices of a network with time-varying edges. In this setting, the learner is restricted to local movement, selecting only its current node or an immediate nei...

📖 Read original article


194. TabPFN beyond Tabular Data: Calibration and Accuracy on Multimodal Embeddings ​

Author: Jingxiang Zhang, Lujia Zhong, Zijie Zhu, Shuo Huang, Yuang Xu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11007v3 Announce Type: replace Abstract: Few-shot multimodal classification commonly attaches a lightweight head, such as $k$-nearest neighbors, logistic regression, or a linear SVM, to a frozen pretrained encoder. Although computationally efficient, these heads can produce poorly calibra...

📖 Read original article


195. Certified-Gap Dual-Price Policies for Real-Time Truckload Bid Acceptance with Relocating, Clock-Constrained Resources ​

Author: Aswin Chandrasekaran
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC

arXiv:2607.16891v2 Announce Type: replace Abstract: A truckload carrier must accept or reject each load tender within seconds. The decision depends on fleet state, hours-of-service (HOS) clocks, and appointment windows. We model this as a weakly coupled dynamic program in which the resources relocat...

📖 Read original article


196. Reliability Scales Inversely: Hallucinations Snowball Faster in Bigger Language Models ​

Author: Kushal Chakrabarti
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.18292v3 Announce Type: replace Abstract: Bigger language models are less reliable. Across three families, three benchmarks and six rungs, including in-the-wild chat logs, scaling closes the start-of-response knowledge gap up to $7\times$ while within-response knowledge degradation grows u...

📖 Read original article


197. Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary ​

Author: Jan Kirin
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.18553v3 Announce Type: replace Abstract: Can a language model read the quality of its ongoing computation, and can an external intervention turn that readout into better outcomes? We test both questions in a frozen 2.6B looped transformer, Ouro-RLTT. On GSM8K, a strict pre-answer probe ex...

📖 Read original article


198. Scaling Laws for Classical Machine Learning on Tabular Data: A Benchmark Study ​

Author: Kaihua Ding
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.21866v2 Announce Type: replace Abstract: Prior classical-ML learning-curve work fits power laws to tree, linear, and kernel models on tabular data, but at small scale: typically one curve, one team, a handful of cells. We present a distributed classroom-scale replication: 127 graduate stu...

📖 Read original article


199. Learned Interventions in Lean 4 grind ​

Author: Evan Wang, Simon Chess, Sophie Szeto, Theodore Meek
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.22972v2 Announce Type: replace Abstract: Lean 4's grind tactic combines congruence closure, E-matching, and case-splitting into a single automated solver, and like any such solver, it relies on hand-tuned heuristics to decide what to instantiate and where to case-split. These heuristics a...

📖 Read original article


200. Variance-Preserving Orthogonal Selection (VPOS): Greedy Feature Selection via Orthogonal Deflation in PCA Loading Space ​

Author: Baran Koseoglu, Berrin Yanikoglu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.23198v2 Announce Type: replace Abstract: We propose Variance-Preserving Orthogonal Selection (VPOS), a greedy framework for unsupervised feature selection that operates in the weighted PCA loading space. After each selection, VPOS projects out the chosen feature's variance direction via n...

📖 Read original article


201. Directional Influence Function: Estimating Training Data Influence in Constrained Learning ​

Author: Xin Wang (Jeff), R. Tyrrell Rockafellar (Jeff), Xuegang (Jeff), Ban
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.23388v3 Announce Type: replace Abstract: As constrained learning becomes increasingly common, models are trained under explicit feasibility requirements to enforce fairness, safety, robustness, regulariza- tion, and physics or logic constraints. Understanding how training samples in- flue...

📖 Read original article


202. A Comparison of Data Augmentation Methods for Training Deep Neural Networks on Synthetic Aperture Sonar ​

Author: C. J. Moore, Gregory D. Vetaw, Jordan Malof
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.23770v3 Announce Type: replace Abstract: In this work we study Automatic Target Recognition (ATR) for Synthetic Aperture Sonar (SAS) data with a focus on deep neural networks (DNNs). The main challenge in training DNNs for SAS-ATR arises from the limited quantity of labeled target example...

📖 Read original article


203. Covariance Last-Layer Ensembles: Function-Space Diversity for Efficient Uncertainty Quantification ​

Author: H. Martin Gillis, Isaac Xu, Gabriel Spadon, Thomas Trappenberg
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.23856v2 Announce Type: replace Abstract: A Last-Layer Ensemble (LLE), $K$ linear units on one shared frozen feature map, is an efficient single-pass approach to the disagreement-based epistemic uncertainty for out-of-distribution (OOD) detection. Its weakness is that members share the bac...

📖 Read original article


204. From Machine Learning to Large-Scale EO Products: Best Practices for Making Maps ​

Author: Ghjulia Sialelli, Robin Young, Yuchang Jiang, Cesar Aybar, Linus Scheibenreif, Damien Robert, Clemens Mosig, Adam J. Stewart, Jan D. Wegner, Aleksis Pirinen, Olof Mogren, Konrad Schindler
Published: 7/29/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.24532v2 Announce Type: replace Abstract: Recent years have seen a rapid expansion in the production of large-scale geospatial maps derived from Earth observation (EO) data, driven largely by advances in machine learning (ML) and large computing infrastructure. Although the barrier to gene...

📖 Read original article


205. A2D2: Audi Autonomous Driving Dataset ​

Author: Jakob Geyer, Yohannes Kassahun, Mentar Mahmudi, Xavier Ricou, Rupesh Durgesh, Andrew S. Chung, Lorenz Hauswald, Viet Hoang Pham, Maximilian M"uhlegg, Sebastian Dorn, Tiffany Fernandez, Martin J"anicke, Sudesh Mirashi, Chiragkumar Savani, Martin Sturm, Oleksandr Vorobiov, Martin Oelker, Sebastian Garreis, Peter Schuberth
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV

arXiv:2004.06320v2 Announce Type: replace-cross Abstract: Research in machine learning, mobile robotics, and autonomous driving is accelerated by the availability of high quality annotated data. To this end, we release the Audi Autonomous Driving Dataset (A2D2). Our dataset consists of simultaneousl...

📖 Read original article


206. Keypoint-Guided Optimal Transport: Models, Algorithms, and Applications ​

Author: Xiang Gu, Yucheng Yang, Wei Zeng, Jian Sun, Zongben Xu
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2303.13102v2 Announce Type: replace-cross Abstract: Existing Optimal Transport (OT) methods mainly derive the optimal transport plan/matching under the criterion of transport cost/distance minimization, which may cause incorrect matching in some cases. In real applications, annotating a few ma...

📖 Read original article


207. FFNet: MetaMixer-based Efficient Convolutional Mixer Design ​

Author: Seokju Yun, Dongheon Lee, Youngmin Ro
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2406.02021v3 Announce Type: replace-cross Abstract: Transformer, composed of self-attention and Feed-Forward Network, has revolutionized the landscape of network design across various vision tasks. While self-attention is extensively explored as a key factor in performance, FFN has received li...

📖 Read original article


208. Representation Capacity-Matched QNN-SNN Twin Construction for Rate-Encoded SNNs ​

Author: Zhanglu Yan, Zhenyu Bai, Kaiwen Tang, Weng-Fai Wong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG

arXiv:2409.08290v5 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs) promise higher energy efficiency over conventional Quantized Artificial Neural Networks (QNNs) due to their event-driven, spike-based computation. However, prevailing energy evaluations often oversimplify, focus...

📖 Read original article


209. A context-adaptive policy framework for robust and reactive robotic manipulation via uncertainty-aware imitation learning ​

Author: Tim R. Winter, Leonard Kl"upfel, Ashok M. Sundaram, Werner Friedl, Maximo A. Roa, Freek Stulp, Jo~ao Silv'erio
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2410.24035v2 Announce Type: replace-cross Abstract: Generating robust and reactive manipulation strategies that can adapt to changing context information is a challenging task in robotics. Over the years, Learning from Demonstration (LfD) has emerged as an intuitive and effective solution for ...

📖 Read original article


210. On the Convergence Analysis of Muon ​

Author: Wei Shen, Ruichuan Huang, Minhui Huang, Cong Shen, Jiawei Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT, math.OC

arXiv:2505.23737v3 Announce Type: replace-cross Abstract: The majority of parameters in neural networks are naturally represented as matrices. However, most commonly used optimizers treat these matrix parameters as flattened vectors during optimization, potentially overlooking their inherent structu...

📖 Read original article


211. TaylorPODA: A Taylor Expansion-Based Method to Improve Post-Hoc Attributions for Opaque Models ​

Author: Yuchi Tang, I~naki Esnaola, George Panoutsos
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2507.10643v4 Announce Type: replace-cross Abstract: Post-hoc model-agnostic local attribution (LA) methods have been widely adopted to explain opaque AI models by quantifying feature-wise contributions. However, many existing methods rely on heuristic or only partially justified attribution me...

📖 Read original article


212. Machine Learning the H-theorem ​

Author: Ruben Lier
Published: 7/29/2026, 4:00:00 AM
Categories: cond-mat.stat-mech, cs.LG

arXiv:2508.14003v4 Announce Type: replace-cross Abstract: The H-theorem provides a microscopic foundation for the Second Law of Thermodynamics and therefore occupies a central place in statistical physics. At the same time, its relation to microscopic reversibility has remained conceptually subtle. ...

📖 Read original article


213. Machine Intelligence on the Edge: Interpretable Cardiac Pattern Localisation Using Reinforcement Learning ​

Author: Haozhe Tian, Qiyu Rao, Nina Moutonnet, Pietro Ferraro, Danilo Mandic
Published: 7/29/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2508.21652v2 Announce Type: replace-cross Abstract: Matched filters are widely used to localise signal patterns due to their high efficiency and interpretability. However, their effectiveness deteriorates for low signal-to-noise ratio (SNR) signals, such as those recorded on edge devices, wher...

📖 Read original article


214. Building Large-Scale English-Romanian Literary Translation Resources with Open Models ​

Author: Mihai Nadas, Laura Diosan, Andreea Tomescu, Andrei Piscoran
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2509.07829v4 Announce Type: replace-cross Abstract: Literary translation has recently gained attention as a distinct and complex task in machine translation research, yet translation by small open models remains an open problem, particularly for low-resource languages such as Romanian. We intr...

📖 Read original article


215. Extreme Event Aware ($\eta$-) Learning ​

Author: Kai Chang, Themistoklis P. Sapsis
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.DS, math.NA

arXiv:2510.19161v2 Announce Type: replace-cross Abstract: Quantifying and predicting rare and extreme events is challenging because such events are infrequent, severe, and expensive to simulate. Existing data-driven methods often require multiple extremes in the training data or sampling process, le...

📖 Read original article


216. A Support-Set Algorithm for Optimization Problems with Nonnegative and Orthogonal Constraints ​

Author: Lei Wang, Xin Liu, Xiaojun Chen
Published: 7/29/2026, 4:00:00 AM
Categories: math.OC, cs.LG, stat.ML

arXiv:2511.03443v2 Announce Type: replace-cross Abstract: In this paper, we investigate optimization problems with nonnegative and orthogonal constraints, where any feasible matrix of size $n \times p$ exhibits a sparsity pattern such that each row accommodates at most one nonzero entry. Our analysi...

📖 Read original article


217. DeepVRegulome: DNABERT-based deep-learning framework for predicting the functional impact of short genomic variants on the human regulome ​

Author: Pratik Dutta, Matthew Obusan, Rekha Sathian, Max Chao, Pallavi Surana, Nimisha Papineni, Yanrong Ji, Zhihan Zhou, Han Liu, Alisa Yurovsky, Ramana V Davuluri
Published: 7/29/2026, 4:00:00 AM
Categories: q-bio.GN, cs.AI, cs.LG

arXiv:2511.09026v2 Announce Type: replace-cross Abstract: Whole-genome sequencing (WGS) has revealed numerous non-coding short variants whose functional impacts remain poorly understood. Despite recent advances in deep-learning genomic approaches, accurately predicting and prioritizing clinically re...

📖 Read original article


218. Online learning of subgrid-scale models for quasi-geostrophic turbulence in planetary interiors ​

Author: Hugo Frezat, Thomas Gastine, Alexandre Fournier
Published: 7/29/2026, 4:00:00 AM
Categories: physics.flu-dyn, astro-ph.EP, cs.LG

arXiv:2511.14581v2 Announce Type: replace-cross Abstract: Machine learning approaches to subgrid-scale (SGS) modelling are now well established in atmospheric and oceanic applications. Among these, online end-to-end learning, where the differentiable solver participates in the training, has shown pa...

📖 Read original article


219. HeatACO: A Heatmap-Guided Max--Min Ant System for Large-Scale Travelling Salesman Problems ​

Author: Bo-Cheng Lin, Yi Mei, Mengjie Zhang
Published: 7/29/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2601.19041v2 Announce Type: replace-cross Abstract: Non-autoregressive neural solvers predict an edge-confidence heatmap for the Travelling Salesman Problem (TSP) in one forward pass, but a decoder must still produce a feasible Hamiltonian cycle. As instance size grows, this stage must reconci...

📖 Read original article


220. Automated Modernization of Machine Learning Engineering Notebooks for Reproducibility ​

Author: Bihui Jin, Kaiyuan Wang, Pengyu Nie
Published: 7/29/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2602.07195v2 Announce Type: replace-cross Abstract: Interactive computational notebooks (e.g., Jupyter notebooks) are widely used in machine learning engineering (MLE) to program and share end-to-end pipelines, from data preparation to model training and evaluation. However, environmental eros...

📖 Read original article


221. Gradient Networks for Universal Magnetic Modeling of Synchronous Machines ​

Author: Junyi Li, Tim Foissner, Floran Martin, Antti Piippo, Marko Hinkkanen
Published: 7/29/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2602.14947v3 Announce Type: replace-cross Abstract: This paper presents a physics-constrained neural network framework for dynamic modeling of saturable synchronous machines, including spatial harmonics. The proposed architecture embeds gradient networks directly into the fundamental machine e...

📖 Read original article


222. Generalization from Low- to Moderate-Resolution Spectra with Neural Networks for Stellar Parameter Estimation: A Case Study with DESI ​

Author: Xiaosheng Zhao, Yuan-Sen Ting, Rosemary F. G. Wyse, Alexander S. Szalay, Yang Huang, L'aszl'o Dobos, Tam'as Budav'ari, Viska Wei
Published: 7/29/2026, 4:00:00 AM
Categories: astro-ph.SR, astro-ph.GA, cs.LG

arXiv:2602.15021v2 Announce Type: replace-cross Abstract: Cross-survey generalization is a critical challenge in stellar spectral analysis, particularly in cases such as transferring from low- to moderate-resolution surveys. We investigate this problem using pre-trained models, focusing on simple ne...

📖 Read original article


223. Emotion Collider: Dual Hyperbolic Mirror Manifolds for Sentiment Recovery via Anti Emotion Reflection ​

Author: Rong Fu, Ziming Wang, Shuo Yin, Kun Liu, Xianda Li, Simon Fong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.MM, cs.CL, cs.LG

arXiv:2602.16161v5 Announce Type: replace-cross Abstract: Emotional expression underpins natural communication and effective human-computer interaction. We present Emotion Collider (EC-Net), a hyperbolic hypergraph framework for multimodal emotion and sentiment modeling. EC-Net represents modality h...

📖 Read original article


224. LiveGraph: Active-Structure Neural Re-ranking for Exercise Recommendation ​

Author: Rong Fu, Zijian Zhang, Jiekai Wu, Kun Liu, Xianda Li, Haoyu Zhao, Yang Li, Yongtai Liu, Ziming Wang, Rui Lu, Simon Fong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2602.17036v5 Announce Type: replace-cross Abstract: The continuous expansion of digital learning environments has catalyzed the demand for intelligent systems capable of providing personalized educational content. While current exercise recommendation frameworks have made significant strides, ...

📖 Read original article


225. CityGuard: Graph-Aware Private Descriptors for Bias-Resilient Identity Search Across Urban Cameras ​

Author: Rong Fu, Yibo Meng, Jia Yee Tan, Rui Lu, Jiekai Wu, Simon Fong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2602.18047v5 Announce Type: replace-cross Abstract: City-scale person re-identification across distributed cameras must handle severe appearance changes from viewpoint, occlusion, and domain shift while complying with data protection rules that prevent sharing raw imagery. We introduce CityGua...

📖 Read original article


226. Standard Transformers Achieve the Minimax Rate in Nonparametric Regression with $C^{s,\lambda}$ Targets ​

Author: Yanming Lai, Defeng Sun
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT

arXiv:2602.20555v2 Announce Type: replace-cross Abstract: The tremendous success of Transformer models in fields such as large language models and computer vision necessitates a rigorous theoretical investigation. To the best of our knowledge, this paper is the first work proving that standard Trans...

📖 Read original article


227. FlashEvaluator: Expanding Search Space with Parallel Sequence-Level Evaluation ​

Author: Chao Feng, Yuanhao Pu, Chenghao Zhang, Shanqi Liu, Shuchang Liu, Xiang Li, Chunjie Chen, Kaiqiao Zhan
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.CL, cs.LG

arXiv:2603.02565v2 Announce Type: replace-cross Abstract: The Generator-Evaluator (G-E) framework generates K candidate sequences and uses an evaluator to select the highest-scoring one, which is widely used in recommender systems (RecSys) and natural language processing (NLP). Existing evaluators c...

📖 Read original article


228. SwiftGS: Episodic Priors for Immediate Satellite Surface Recovery ​

Author: Rong Fu, Jiekai Wu, Xiaowen Ma, Shiyin Lin, Kangan Qian, Chuang Liu, Simon James Fong
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2603.18634v4 Announce Type: replace-cross Abstract: Rapid, large-scale 3D reconstruction from multi-date satellite imagery is vital for environmental monitoring, urban planning, and disaster response, yet remains difficult due to illumination changes, sensor heterogeneity, and the cost of per-...

📖 Read original article


229. RankFormer: A Propose-then-Select Transformer for Multi-Agent Multimodal Trajectory Prediction ​

Author: Diyi Liu, Zihan Niu, Tu Xu, Xingchen Zhang, Lishan Sun
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2604.07126v2 Announce Type: replace-cross Abstract: Predicting vehicle trajectories plays an important role in autonomous driving, transportation safety analysis, traffic operations, etc. Although many deep learning algorithms are devised to predict future vehicle trajectories, the vehicle tra...

📖 Read original article


230. SCOPE-FE: Structured Control of Operator and Pairwise Exploration for Feature Engineering via Quality-Aware Candidate-Space Reduction ​

Author: Minhee Park, Seongyeon Son, Yonghyun Lee, Eunchan Kim
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2604.27025v2 Announce Type: replace-cross Abstract: Automatic feature engineering can improve predictive performance on tabular data by generating diverse feature transformations. However, the candidate space induced by combinations of input features and operators grows rapidly with dimensiona...

📖 Read original article


231. FREPix: Frequency-Heterogeneous Flow Matching for Pixel-Space Image Generation ​

Author: Mingfeng Lin, Jiakun Chen, Liang Han, Liqiang Nie
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2605.06421v2 Announce Type: replace-cross Abstract: Pixel-space diffusion has re-emerged as a promising alternative to latent-space generation because it avoids the representation bottleneck introduced by VAEs. Yet most existing methods still treat image generation as a frequency-homogeneous p...

📖 Read original article


232. Spatially Grounded Concept Bottleneck Models via Part-Factorized Attention ​

Author: Dhanesh Ramachandram
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2606.04364v3 Announce Type: replace-cross Abstract: Concept bottleneck models (CBMs) predict a layer of human-named attributes before predicting a class, which makes their decisions auditable. On fine-grained recognition tasks, though, the concept heads are usually free to attend anywhere in t...

📖 Read original article


233. Multi-Fidelity Learning with Shallow Recurrent Decoders for Multi-Physics Applications ​

Author: Stefano Riva, Carolina Introini, J. Nathan Kutz, Antonio Cammi
Published: 7/29/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG

arXiv:2606.05202v2 Announce Type: replace-cross Abstract: In reactor physics, neutronics and multi-physics phenomena can be modelled at different fidelity levels. High-fidelity models based on the Boltzmann transport equation, multi-group diffusion, or computational fluid dynamics are computationall...

📖 Read original article


234. Wall Shear Stress Reconstruction from Concentration: Differentiable Physics and Physics-Informed Neural Networks ​

Author: Mahmoud Elhadidy, Siva Viknesh, Roshan M. D'Souza, Amirhossein Arzani
Published: 7/29/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG, physics.comp-ph

arXiv:2606.06313v2 Announce Type: replace-cross Abstract: Wall shear stress (WSS) governs near-wall transport dynamics and is a key hemodynamic indicator in cardiovascular flows, yet remains difficult to infer accurately due to the need for precise computation of near-wall velocity gradients. Passiv...

📖 Read original article


235. Practical Quantum Advantage before Fault Tolerance via Quantum-Informed Machine Learning ​

Author: Maida Wang, Xiao Xue, Minh Chung, Peter V. Coveney
Published: 7/29/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, physics.flu-dyn

arXiv:2606.13422v4 Announce Type: replace-cross Abstract: Early quantum devices can deliver a practical advantage before fault tolerance. The role we identify is a statistical module within a classical scientific workflow: a compressed memory with a collective two-copy readout, evaluated against a v...

📖 Read original article


236. Optimal Ansatz-free Hamiltonian Learning In Situ ​

Author: Taiqi Zhou, Weiyuan Gong
Published: 7/29/2026, 4:00:00 AM
Categories: quant-ph, cs.IT, cs.LG, math.IT

arXiv:2606.19486v2 Announce Type: replace-cross Abstract: Characterizing the features of a Hamiltonian that governs a quantum system serves as a fundamental subroutine of quantum device calibration, signal sensing, and error correction. Recent works have proposed protocols achieving the optimal Heis...

📖 Read original article


237. Token Factory: Efficiently Integrating Diverse Signals into Large Recommendation Models ​

Author: Xilun Chen, Shao-Chuan Wang, Baykal Cakici, Lukasz Heldt, Lichan Hong, Raghu Keshavan, Aniruddh Nath, Li Wei, Xinyang Yi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2606.19635v3 Announce Type: replace-cross Abstract: Large Recommendation Models (LRMs) have demonstrated promising capabilities in industry-scale recommendation tasks. However, holistically integrating traditional signals into these transformer-based architectures effectively and efficiently r...

📖 Read original article


238. RoboMME-Interference: Benchmarking Robot Memory Under Interference ​

Author: Soumil Rathi
Published: 7/29/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2606.22338v2 Announce Type: replace-cross Abstract: Robots deployed in realistic settings will accumulate experience across many sessions and tasks over their deployment. The robot's tasks may often require it to remember information from multiple sessions ago, making long-context robot memory...

📖 Read original article


239. Fitted Occupancy-Ratio Evaluation without Bellman Completeness ​

Author: Lars van der Laan, Nathan Kallus
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.05375v2 Announce Type: replace-cross Abstract: Occupancy ratios correct distribution shift in offline reinforcement learning and are central to off-policy evaluation. Existing primal-dual and minimax methods typically estimate these ratios by enforcing occupancy-balance moments over a cri...

📖 Read original article


240. Contextual Procurement Auctions with Bandit Learning ​

Author: Yiling Chen, Shi Feng, Sadie Zhao
Published: 7/29/2026, 4:00:00 AM
Categories: cs.GT, cs.LG

arXiv:2607.05813v3 Announce Type: replace-cross Abstract: We study repeated procurement auctions in which producers have private costs and the platform must learn the context-dependent value of selecting each producer. We evaluate performance by welfare regret: the cumulative loss in total surplus r...

📖 Read original article


241. When Prompts Ignore Structure: Graph-Based Attribute Reasoning for Calibrated VLMs ​

Author: Tanay Sodha, Aditya Sharma, Ramya Hebbalaguppe, Vinti Agarwal, Pranav Murthy Yeluripaty
Published: 7/29/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.07395v2 Announce Type: replace-cross Abstract: Reliable confidence estimation remains a key limitation of test-time adaptation in vision-language models (VLMs), where prompt tuning improves zero-shot accuracy but often degrades calibration due to entropy-driven overconfidence. Prior appro...

📖 Read original article


242. Tokenizing Numerical and Embedding Features for LLM RecSys ​

Author: Zhe Xu, Ankit Peshin, Chiyu Zhang, Feng Qi, Johnson Lui, Anil Ramakrishna, Justin Johnson, Carl Hu, Kaushik Rangadurai, Luke Simon
Published: 7/29/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.10016v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used as backbone architectures for recommender systems because of their strong sequence modeling and representation learning capabilities. However, most LLM-based recommenders operate primarily on...

📖 Read original article


243. Input-Aware Dynamic Backdoor Attack Against Quantum Neural Networks ​

Author: Junrui Zhang, Zemin Chen, Lusi Li, Mohammad Ghasemigol, Daniel Takabi, Rui Ning
Published: 7/29/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2607.11843v2 Announce Type: replace-cross Abstract: Quantum Neural Networks (QNNs) are a promising framework for quantum machine learning on near-term quantum devices, but their security risks remain insufficiently understood. Studies have shown that QNNs are vulnerable to backdoor attacks, ye...

📖 Read original article


244. On the Order-Conditional Optimality of Gaffke's Bound ​

Author: George Bissias, Erik Learned-Miller
Published: 7/29/2026, 4:00:00 AM
Categories: math.ST, cs.LG, math.PR, stat.ML, stat.TH

arXiv:2607.22971v2 Announce Type: replace-cross Abstract: Let $X = (X_1, \ldots, X_n)$ be a random vector from any Borel probability law on $\mathbb{R}_+^n$. We revisit the problem of deriving a lower confidence bound (LCB) on a scalar parameter of that law. We recast classical work, beginning with ...

📖 Read original article


245. Distributed Convolutional Rank Regression over Decentralized Networks ​

Author: Chunjing Li, Tiange Zhao, Xiaohui Yuan
Published: 7/29/2026, 4:00:00 AM
Categories: stat.ME, cs.LG

arXiv:2607.23639v2 Announce Type: replace-cross Abstract: This paper studies convolution rank regression (CRR) over decentralized distributed learning networks. We propose a novel decentralized CRR framework, in which estimators are obtained by solving consensus-constrained optimization with kernel-...

📖 Read original article