arXiv cs.LG - 2026-07-21 ​
481 items collected.
1. Reinforcement Learning-Guided NSGA-II Enhanced with Gray Relational Coefficient for Multi-Objective Optimization: Application to NASDAQ Portfolio Optimization ​
Author: Zhiyuan Wang, Qinxu Ding, Ding Ding, Siying Zhu, Jing Ren, Yue Wang, Chong Hui Tan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, econ.GN, q-fin.EC
arXiv:2607.16194v1 Announce Type: new Abstract: In modern financial markets, decision-makers increasingly rely on quantitative methods to navigate complex trade-offs among multiple, often conflicting objectives. This paper addresses constrained multi-objective optimization (MOO) with an application ...
2. DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth ​
Author: Zihan Xu, Puzhen Wu, Lawrence Chun Man Lau, Wei Liu, Sirui Li, Yifan Peng, Yihao Ding
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.16203v1 Announce Type: new Abstract: Document parsing is a foundational step for document understanding tasks such as visual question answering and key information extraction, as it transforms unstructured scanned images into structured representations by extracting textual, visual, and l...
3. Fully-sensorized smart-eyewear platform for on-device Machine Learning ​
Author: Andrea Giudici, Christian Veronesi, Pietro Bartoli, Mario Cali`o, Aurelio Teliti, Giacomo Gervasoni, Diana Trojaniello, Franco Zappa
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP
arXiv:2607.16222v1 Announce Type: new Abstract: This paper presents ARGO, a smart eyewear platform designed to bridge ergonomic comfort, high computational throughput, and energy efficiency. Unlike cloud-dependent solutions, ARGO leverages the STM32N6 microcontroller and its integrated Neural Proces...
4. LLM Unlearning for Cyber Defense: A Survey on Methods, Challenges, and Emerging Threats ​
Author: Ruppikha Sree Shankar, Abhishek Bhardwaj, Arnav Doshi, Anusri Nagarajan, Troy Paulus Asia, Saptarshi Sengupta
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2607.16227v1 Announce Type: new Abstract: LLMs are increasingly deployed in security-critical systems across healthcare, finance, education, and decision support, yet their inability to forget creates serious cybersecurity, privacy, and safety risks. Sensitive personal information, copyrighted...
5. Operator-Aware Mixed-Precision Tolerance Calibration for Tensor Kernels ​
Author: Dipankar Sarkar
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.MS
arXiv:2607.16228v2 Announce Type: new Abstract: Most tensor-kernel correctness tests go through a fixed-shape all close-style check with hand-picked absolute and relative tolerances. The thresholds are copied across the corpus and rarely revisited. We mine the element-wise error distribution of ever...
6. RouteCost: A Production-Inspired Multi-Stage Framework for Pre-Order Shipping Cost Estimation in E-Commerce ​
Author: Xianling Zeng, Zihan Yu, Sichen Zhao, Yalun Qi, Zhiming Xue
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16230v1 Announce Type: new Abstract: Accurate pre-order shipping cost estimation is important in e-commerce because it affects price presentation, margin planning, and conversion. In practice, shipping cost is shaped not only by distance but also by destination demand mix, billable weight...
7. Orthogonal Gradient Constraints Shape Noisy-Label Memorization Dynamics ​
Author: Richard Mai
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16231v1 Announce Type: new Abstract: Modern neural networks can fit corrupted training labels, making noisy-label learning a useful setting for studying memorization-driven overfitting. Most regularization methods modify the objective, architecture, or data distribution; here we instead s...
8. From Weights to Words: Expressing and Editing Preference Model Inferences in Natural Language ​
Author: Zachary Wojtowicz, Ayush Nayak, Jacob Andreas
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.HC
arXiv:2607.16232v1 Announce Type: new Abstract: The growing use of statistical learning algorithms to infer human preferences from high-dimensional choice data runs up against a fundamental challenge: choice alternatives typically differ in many ways simultaneously, so it is generally unclear which ...
9. Token-Level Cross-Modal Transformer with Contrastive Multi-Task Learning for Breast Cancer Subtype Classification and Survival Prediction ​
Author: Suxing Liu Byungwon Min
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16233v1 Announce Type: new Abstract: Integrating heterogeneous genomic and clinical modalities for joint cancer subtype classification and survival prediction remains a key challenge in precision oncology. Existing approaches suffer from three limitations: (1) they treat each modality as ...
10. HantaWatch: Federated Learning for Hantavirus Genomic Surveillance ​
Author: Shanika Iroshi Nanayakkara, Shiva Raj Pokhrel
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16234v1 Announce Type: new Abstract: Hantavirus genomic surveillance is limited by the distribution of sequence data, non-IID source heterogeneity, and constrained expert-review capacity. We propose HantaWatch, a federated learning framework that enables laboratories and surveillance site...
11. OpenMHC: Accelerating the Science of Wearable Foundation Models ​
Author: Narayan Schuetz, Yuze Bai, Lianggang Pan, Edgar Eggert, Favour Nerrise, Juan Delgado-SanMartin, Max Rosenblattl, Milana Gurbanova, Mohammad Asadi, Anders Johnson, Paul Schmiedmayer, Dennis Wang, Allan Lawrie, Daniel Seung Kim, Xin Liu, Akshay Paruchuri, Ehsan Adeli, Euan Ashley, Kelly W. Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16235v2 Announce Type: new Abstract: Mobile and wearable devices offer an unprecedented opportunity for continuous, passive health monitoring and active health coaching. However, the largest wearable datasets are not publicly available for research, and leading wearable foundation models ...
12. The Failures of Marginal Influence-Based Attribution Methods for Global Time Series Explanations ​
Author: Amadeo Tunyi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16236v1 Announce Type: new Abstract: Explainability methods for time series models predominantly produce flat attribution scores: they quantify the direct influence of a feature at a timestamp by a scalar. We prove that the dominant failure mode of such methods is not the scalar format it...
13. Quantizing Recursive Reasoning Models ​
Author: Thorir Mar Ingolfsson, Wajeeha Tahir, Anna Tegon, Lionnus Kesting, Gamze .Islamo\u{g}lu, Luca Benini
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16237v1 Announce Type: new Abstract: Recursive reasoning models solve hard puzzles by applying compact, weight-tied blocks over many refinement steps. Because these blocks are reused many times, quantizing them creates a unique dynamical problem: the quantization error is incurred at ever...
14. Diffusion-corrected Autoregressive Fourier Neural Operator for Droplet Evolution Prediction ​
Author: Jinghao Cao, Minsung Kang, Hongyue Sun, Chi Zhou, Jihoon Chung, Xubo Yue, Sanchoy Das, Bo Shen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, physics.comp-ph, physics.flu-dyn
arXiv:2607.16238v1 Announce Type: new Abstract: Predicting droplet evolution in material jetting, or Inkjet Printing (IJP), is essential for maintaining printing quality. However, long-horizon forecasts remain challenging due to error accumulation and the complex coupling of process variables. In th...
15. BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges ​
Author: Lei Shi, Anlan Zhang, Rita Lyu, Zhengmian Hu, Tong Yu, David Arbour, Avi Feller, Saayan Mitra, Ritwik Sinha
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, stat.AP
arXiv:2607.16239v1 Announce Type: new Abstract: AI judges offer a scalable, low-cost alternative to human evaluation, but their outputs can be biased relative to human preferences and highly item-dependent, varying across judges, tasks, and domains. When uncalibrated AI evaluations are used for mode...
16. Normalized Rewards for Preference Optimization ​
Author: Shawn Im, Federico Danieli, Skyler Seto, Barry-John Theobald, Katherine Metcalf
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16240v1 Announce Type: new Abstract: Direct Alignment Algorithms (DAAs) such as DPO have become a common way to post-train and align LLMs with human preferences. However, DAAs have been observed to over-optimize their implicit reward model and decrease the likelihood of preferred response...
17. KernelBench-Verified: Do LLM-Generated Kernels Actually Beat PyTorch? ​
Author: Yunxiang Zhang (Xiangjun), Ping Yu (Xiangjun), Jianyu Wang (Xiangjun), Max (Xiangjun), Fan, Julian Reed, Azalia Mirhoseini, Will Su
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16241v1 Announce Type: new Abstract: Recent large language models (LLMs) can generate custom CUDA kernels that appear to outperform PyTorch on benchmarks such as KernelBench. Building upon this foundational framework, we demonstrate that frontier models frequently engage in reward hacking...
18. TRACE: Trajectory-Based Safety Patch Learning for LLM Post-Training Realignment ​
Author: Changyue Li, Jiaming He, Youliang Yuan, Jialin Wu, Boxi Yu, Zhicong Huang, Pinjia He
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2607.16242v1 Announce Type: new Abstract: Fine-Tuning-as-a-Service (FTaaS) platforms let users train large language models (LLMs) on customized tasks, but this pipeline could erode models' safety alignment. In practice, service providers need to recover models' safety without re-running full a...
19. RobustMAD: Evaluating Real-World Robustness of Multimodal Small Language Models for Deployable Anomaly Detection Assistants ​
Author: Anushiya Arunan, Xin Li, Yan Qin, U-Xuan Tan, Nhu Khue Vuong, Xiaoli Li, Chau Yuen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16243v1 Announce Type: new Abstract: Multimodal industrial anomaly inspection assistants are a critical component of next-generation smart factories, enabling interactive vision-language-based querying. However, multimodal large language models remain impractical for on-site deployment du...
20. CIGPO: Contextual Information-Gain Policy Optimization for Multi-Turn Evidence-Reading LLM Agents ​
Author: Hao Dou
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.16244v1 Announce Type: new Abstract: Training multi-turn evidence-reading agents with outcome-only reinforcement learning is unstable because intermediate turns receive little direct credit. In HotpotQA experiments with Qwen2.5-3B-Instruct, GRPO initially improves (standard F1 0.430) but ...
21. Learning Structural Manipulability in Gate-Level Netlists Using Graph Neural Networks ​
Author: Rupesh Raj Karn, Ozgur Sinanoglu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16245v1 Announce Type: new Abstract: Gate-level netlists exhibit intrinsic structural properties that influence signal propagation independently of functional simulation. We define a topology-driven structural manipulability score that characterizes node-level structural flexibility using...
22. Let the Data Decide: Supervision Analysis, Capability Trade-offs, and Adaptive Objective Routing in Continued Pre-Training via Off-Policy Distillation ​
Author: Jiangan Yuan, Zhixuan Li, Han Xu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16246v1 Announce Type: new Abstract: Off-policy distillation is now central to large language model pre-training, yet how training data, objective parameterization, and model capabilities interact remains poorly characterized. We studies top-$k$-truncated, temperature-scaled off-policy di...
23. Self-Evolving Just-In-Time Memory for Proactive Embodied Safety ​
Author: Bingrui Sima, Lizhong Wang, Xiaoya Lu, Kun He, Xiao Yang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.16247v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have empowered embodied agents to execute complex household tasks, they struggle to proactively handle dynamically emerging hazards during closed-loop interactions. Existing safety approaches often rely on runtime gu...
24. High-accuracy Low-Bit KV-Cache Quantization via Local Distribution Restoration ​
Author: Gradwell Dzikanyanga, Yanqi Pan, Weihao Yang, Donglei Wu, Wen Xia, Hao Huang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16248v1 Announce Type: new Abstract: Long-context large language model inference relies on the KV cache to avoid redundant attention computation, but incurs high memory and bandwidth overheads. Low-bit KV-cache quantization reduces this cost, yet it severely degrade quality; particularly,...
25. Benchmarking Machine Learning Models for Multi-Omics-Based Breast Cancer Prediction ​
Author: Priyanka Paudel, Madan Baduwal
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16250v1 Announce Type: new Abstract: Estrogen Receptor (ER) status is a critical biomarker in breast cancer diagnosis, prognosis, and treatment selection. Recent advances in high-throughput sequencing technologies have enabled the generation of multi-omics datasets that provide complement...
26. Learning Spatio-Temporal Foundation Models from Pure Synthetic Data ​
Author: Yutong Feng, Shiyuan Piao, Yutong Xia, Xu Liu, Wenqi Fan, Fugee Tsung, See-Kiong Ng, Yuxuan Liang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16251v1 Announce Type: new Abstract: Spatio-Temporal Foundation Models (STFMs) aim to learn generalizable representations of complex dynamical systems across space and time. However, existing approaches suffer from distributional bias in real-world pre-training data, structural bottleneck...
27. SOS-LoRA: Static Orthogonal-Subspace Low-Rank Adaptation with Fixed Multi-Scale Scaling ​
Author: Yupeng Chang, Yuan Wu, Yi Chang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16252v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is a widely used parameter-efficient fine-tuning (PEFT) method for large language models. Under a fixed rank budget, LoRA parameterizes each adapted weight through a single low-dimensional input-side pathway, which may couple...
28. Comprehensive Evaluation of Machine Learning for Type 2 Diabetes Risk Prediction: Large-Scale External Validation and Fairness Analysis ​
Author: Rajveer Singh Pall, Sameer Yadav, Siddharth Bhalerao, Sourabh Sahu, Ritu Ahluwalia, Bhaskar Awadhiya
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16253v1 Announce Type: new Abstract: Machine learning-based Type 2 diabetes risk prediction models obtain good internal validation results but lose effectiveness in real-world applications due to deficient external testing and fairness assessment. We developed a multi-dimensional framewor...
29. More Than Memory: Task-Conditioned Signed FFN Writes in Long-Context Retrieval ​
Author: Zhibo Yang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16254v1 Announce Type: new Abstract: FFNs are often treated as parametric memories. In long-context retrieval, however, the sharper question is not only what they store, but whether their native residual writes push the current retrieval state toward or away from the correct answer. We te...
30. Feature Generation Using LLMs: An Evolutionary Algorithm Approach ​
Author: Aria Nourbakhsh, Beno^it Alcaraz, Christoph Schommer
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.16255v1 Announce Type: new Abstract: A crucial step in machine learning pipelines is to present each entity with features or attributes that are representative of the characteristics of the processed entities. Feature engineering is an important step in finding a relation among attributes...
31. Discovery by Dreaming: Cross-Domain Recombination in Artificial Memory ​
Author: Oliver Zahn, James Evans, David Eagleman
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR, cs.NE
arXiv:2607.16256v2 Announce Type: new Abstract: Dreams splice together people, places, and times that never met. Neuroscience suggests this recombination is not noise, but a function driving insight and creative discovery. This reframes memory consolidation: rather than merely defending against forg...
32. From Outcomes to Actions: Leveraging Hindsight for Long-Horizon Language Agent Training ​
Author: Zishang Jiang, Tingyun Li, Jinyi Han, Xinyi Wang, Sihang Jiang, Yizhou Ying, Xiaojun Meng, Jiansheng Wei, Jiaqing Liang, Yanghua Xiao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.16257v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a widely adopted technique for improving large language models (LLMs) on complex tasks. Despite this progress, existing RL methods still face challenges in training agents with longer-horizon interactions. One maj...
33. Neural Controlled Differential Equations for EMT-Level Surrogate Modeling of Grid-Forming Inverters ​
Author: Jiagang Qu, Yong Tao, Dan Wang, Enyi Li, Jingjing Qi, Ding Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16258v1 Announce Type: new Abstract: The application of artificial intelligence methods in power electronic converter modeling is becoming increasingly widespread, but existing applications still face many challenges, such as difficulties in multi-time-scale hybrid analysis and the lack o...
34. Quantifying Ranking Uncertainty in LLM Benchmarks ​
Author: Bitya Neuhof, Yuval Benjamini
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, stat.ML
arXiv:2607.16259v1 Announce Type: new Abstract: Pretrained models are typically ranked on multi-task leaderboards to assess their effectiveness across diverse tasks. Rank confidence intervals were recently introduced as a method to quantify the uncertainty in these rankings by aggregating pairwise h...
35. AdaSurvMamba: Dynamic Fusion and Semantic Scanning for Multimodal Survival Analysis ​
Author: Jialong Zhong, Tingwei Liu, Baokun Yue, Jingjing Li, Yongri Piao, Miao Zhang, Leiye Liu, Jiahong Jiang, Wei Ji, Huchuan Lu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2607.16260v1 Announce Type: new Abstract: Multimodal survival analysis utilizing whole slide images (WSIs) and genomic profiles is fundamental for cancer prognosis. Recently, state-space models like Mamba have emerged as powerful tools for sequence modeling. However, translating this success t...
36. Reducing Per-Sample Harm in Stochastic Optimization ​
Author: Apostolos Avranas
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.16261v1 Announce Type: new Abstract: Modern optimizers combine gradients from the current mini-batch with historical optimization state, such as momentum or adaptive moments. While highly effective, aggregating across the batch and incorporating this history can produce parameter updates ...
37. Autonomous mechanistic discovery of colorectal cancer vulnerabilities via multi-scale AI swarms ​
Author: Christopher Baker, Tianyu Ren, Karen Rafferty, Hui Wang, Simon McDade
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16262v1 Announce Type: new Abstract: The acceleration of automated scientific discovery has been fundamentally bottlenecked by the epistemic gap between the semantic reasoning of large language models (LLMs) and the deterministic physics of mammalian biology. While recent multi-agent fram...
38. Preference-based Antibody Expression Ranking: Scaling with Large-scale Weak Supervision ​
Author: Josh Qixuan Sun, Morteza Babaie, Wenyang Hou, Mark Crowley, David Young
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, q-bio.QM
arXiv:2607.16263v1 Announce Type: new Abstract: Antibody expression ranking is a critical task in antibody design, yet its modelling is severely hindered by the scarcity of labeled experimental data. To address this, we propose a unified preference-based learning framework that integrates scarce qua...
39. PsiLogic: Chaos-Aware Active Cancellation for Adam with a Fair Cross-Domain Benchmark ​
Author: Ali Sultonov
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16268v1 Announce Type: new Abstract: Adaptive optimizers such as Adam and AdamW apply the same update rule regardless of whether training is in a chaotic early phase or near convergence. We introduce PsiLogic, an optimizer that augments Adam with a dynamic Active Cancellation Term gated b...
40. A Predict-then-Correct Loop Based on Few-Shot Continuous Contextual Bandit for Demand Forecasting ​
Author: Zhiwei Lei, Benedict Jun Ma, Ilya Jackson
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16354v1 Announce Type: new Abstract: Retail demand forecasting remains difficult when demand shifts faster than static forecasting models can be retrained, especially in early demand cycles where newly observed labels are sparse. To address this, this study aims to improve adaptive retail...
41. Scaling Limits of Constant-Stepsize SGD at Flat Minima ​
Author: Jingyi Zhang, Cheng Mao, Debankur Mukherjee
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, math.OC, math.PR, stat.ML
arXiv:2607.16384v1 Announce Type: new Abstract: For stochastic gradient descent (SGD) with a constant stepsize $\alpha$, the invariant law of the iterates, centered at a minimizer, describes the behavior of the algorithm over long time horizons. In the strongly convex case, this invariant law has th...
42. Retrieval is Enough: Training-Free Interpretability with a Tool-Using Agent ​
Author: Sriram Balasubramanian, Soheil Feizi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16448v1 Announce Type: new Abstract: Interpretability methods for neural network activations span a wide cost spectrum, from cheap, training-free techniques (such as linear probes, PCA, SVD) to more expensive training-based ones (such as SAEs and activation oracles). Training-based method...
43. EA-RMENet -- Path Loss Prediction in Urban Environments using Deep Learning ​
Author: Jonathan O'Shea (DCU School of Electronic Engineering), Conor Brennan (DCU School of Electronic Engineering)
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16449v1 Announce Type: new Abstract: Accurate path loss prediction is a critical component of wireless network planning. Current path loss prediction methods typically struggle to balance the trade-off between accuracy and computational efficiency. This paper proposes the Efficient Attent...
44. Compact convolutional neural networks for AI-based drone detection system ​
Author: G'abor Farkas, G'abor Fazekas, Karakai Patrik, Andr'as N'emeth, G'abor Farkas
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16455v1 Announce Type: new Abstract: The increasing use of first-person-view drones in modern conflicts has created a demand for compact and reliable detection systems capable of operating in complex electromagnetic environments. These drones continuously transmit video signals through on...
45. K-IPO: Kendall-constrained Importance Preserving Oversampling for Imbalanced Tabular Data ​
Author: Marios Tyrovolas, Argiris Sofotasios, Dimitris Metaxakis, Georgios Mermigkis, George Georgoulas, Panagiotis Hadjidoukas, Chrysostomos Stylios
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16478v1 Announce Type: new Abstract: Oversampling is widely used to address class imbalance in tabular classification, but existing methods can distort the feature importance ranking underlying model explanations. Although recent studies have quantified this distortion by comparing real a...
46. Leakage-Robust Evaluation and Data-Scale Sensitivity of Attention-Enhanced Multi-Task Learning for Joint Fault Diagnosis and Remaining Useful Life Estimation ​
Author: Md Mahamudur Rahaman Shamim, Md. Nuruzzaman, Zannatul Ferdus, Md Rajib Ahmed, Abieer Nwshad Anward, Mohammad Tooneer, Johir Uddin Khan, Khalid Hossen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16493v1 Announce Type: new Abstract: Multi-task deep learning models that jointly perform fault classification and remaining useful life (RUL) regression are increasingly used in predictive maintenance, yet reported performance can be strongly affected by how sliding-window sequences are ...
47. Feedback Attribution and Representation Geometry: Metrics for Comparing Individual and Shared Rewards in MARL ​
Author: Tasha Pais, Richard Higgins
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16524v1 Announce Type: new Abstract: Cooperative multi-agent RL systems routinely use team-averaged rewards, a feedback-attribution choice that gives each agent the team outcome regardless of its individual contribution. We ask whether this leaves a measurable signature, geometric or beha...
48. Hierarchical Domain Generalization ​
Author: Chenxiao Yang, Zhiyuan Li, Shai Ben-David, Nathan Srebro
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16528v1 Announce Type: new Abstract: We study hierarchical domain generalization as a problem of extrapolation from finite observed regions to an entire instance space, replacing i.i.d. sampling with arbitrary domain hierarchies. We show that the central obstruction is not only the comple...
49. Building2Building: A Large Scale Benchmark for Generalizable Real-World Reinforcement Learning ​
Author: Vincent Taboga, Justin Veilleux, Doseok Jang, Anushree Rankawat, Pierre-Luc Bacon
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16534v1 Announce Type: new Abstract: Reinforcement learning (RL) has achieved strong results in control, yet learned policies remain brittle to changes in dynamics, action spaces, observation spaces, or goals, a critical limitation for real-world deployment. Existing benchmarks offer limi...
50. Discrete Ricci Curvature on Protein Contact Graphs for Lightweight Fold Classification ​
Author: Jianru Shen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16553v1 Announce Type: new Abstract: Protein fold classification can be approached via sequence-based representations or structural descriptors, but direct comparisons between lightweight handcrafted descriptors and pretrained protein language model embeddings remain limited. We investiga...
51. Capacity and Redundancy Trade-offs in Multi-Task Learning ​
Author: Asif Khan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT
arXiv:2607.16554v1 Announce Type: new Abstract: In multi-task learning (MTL) negative transfer is often considered as an optimization artifact, but it can also be viewed as a consequence of limited shared capacity and weak task redundancy. We investigate this effect through a Capacity--Redundancy (C...
52. Learning from World Feedback: Why Model Uncertainty Fails as a Risk Signal in Model-Based RL ​
Author: Zhaohui Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16591v1 Announce Type: new Abstract: The RLxF programme argues that learning signals should come from world feedback rather than from internal model proxies. We instantiate this position in safe model-based control and distil it into three concrete design principles. Empirically, across f...
53. Privacy Cost as Equity Input: A Group Fairness Criterion for Differentially Private Machine Learning ​
Author: Rakshit Naidu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR, cs.CY
arXiv:2607.16620v1 Announce Type: new Abstract: Differential privacy (DP) is increasingly deployed to limit membership inference risk in machine-learning systems. Prior work has shown that DP-SGD can widen accuracy disparities across demographic groups, but this framing treats fairness as a purely o...
54. Multimodal Attention-based Deep Learning for Emergency Triage with Electronic Health Records ​
Author: Hazqeel Afyq Athaillah Kamarul Aryffin, Kamarul Aryffin Baharuddin, Mohd Halim Mohd Noor
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16662v1 Announce Type: new Abstract: Accurate emergency triage decision is critical to avoid clinical deterioration, morbidity, and mortality. Machine learning-based triage system involves acquiring the main presenting complaint in text form and assessing vital signs in numerical data, en...
55. CLDRoute: Conditional Latent Diffusion for Routability Map Generation in Physical Design ​
Author: Kiran Thorat, Nicole Meng, Caiwen Ding, Yingjie Lao, Zhijie Jerry Shi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16674v1 Announce Type: new Abstract: Accurate routability estimation during physical design is important for reducing costly post-routing iterations. Prior learning-based methods treat this task as deterministic prediction, mapping placement-stage features to a single congestion or DRC ou...
56. A Framework for Early Sepsis Prediction via Self-Supervised (JEPA) and Federated Representation Learning ​
Author: Umair bin Mansoor, Munaf Rashid, Roomi Naqvi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16681v1 Announce Type: new Abstract: Early sepsis prediction from electronic health records is challenged by irregular sampling, high missingness, and class imbalance. We systematically compare four modeling paradigms -- self-supervised Joint Embedding Predictive Architecture (JEPA) via m...
57. Building a Neural Network from Scratch: Implementation, Evaluation, and Optimization ​
Author: Yuanzhe Jia
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16682v1 Announce Type: new Abstract: The widespread adoption of high-level deep learning libraries, while accelerating model development, has increasingly abstracted away the internal mechanics of neural networks, creating a gap between practical usage and fundamental understanding. To ad...
58. Effects of width-dependent model hyperparameters and $\ell_2$-regularization on the loss landscape of two-layer ReLU networks ​
Author: Haruka Eshima, Makoto Yamada
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16720v1 Announce Type: new Abstract: Understanding deep neural networks remains a central challenge in machine learning. In particular, the theoretical properties of even two-layer ReLU networks, especially in the presence of weight decay, remain poorly understood. To this end, we derive ...
59. Half the Experts, All the Code: One-Shot Domain Pruning of Mixture-of-Experts LLMs for Coding ​
Author: Anik Jha
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16721v1 Announce Type: new Abstract: The strongest open-weight coding models are mixture-of-experts (MoE) networks: most of their size comes from large pools of "expert" subnetworks, of which only a few act on any token. That pool is why these models do not fit on the machines most develo...
60. Graph-Embedded Intuitionistic Fuzzy Broad Learning System: A Multi-view Framework ​
Author: Yogesh Kumar, Manju, Mudasir Ganaie
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16728v1 Announce Type: new Abstract: The Broad Learning System (BLS) has been widely used for data classification and is based on a layer-by-layer feed-forward structure. However, it gives the same importance to all data points, which reduces its effectiveness on real-world datasets with ...
61. BG4Sea: Biogeochemical Seasonal Forecastability via Progressive Information Scaling ​
Author: Gabriela Martinez Balbontin, Anastase Charantonis, Dominique Bereziat, Stefano Ciavatta
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16731v1 Announce Type: new Abstract: Marine biogeochemical forecasting is increasingly important for managing marine ecosystems and the carbon cycle, yet global, seasonal forecast products lag far behind physical oceanography, held back by the complexity of the processes involved and by d...
62. The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models ​
Author: Francesco Karim Vicidomini
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16741v1 Announce Type: new Abstract: B"urger et al. (2024) demonstrated that truth representations in large language models are universal across statement polarity but reside within a multidimensional subspace. We extend that framework along three questions: how the dimensionality of the...
63. Robust Losses from Univariate Base Functions for Noisy-Label Learning ​
Author: Peng Hu, Jianwei Ma
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16768v1 Announce Type: new Abstract: Learning with noisy labels is a fundamental problem in training reliable deep neural networks. Robust loss functions provide a direct and effective way to mitigate the adverse effects of label noise. However, most existing robust losses are designed di...
64. On the Potential of Graph Neural Networks as Metamodels for Supply Chain Optimization: Dataset, Architectures, and Directions ​
Author: Tushar Lone, Neha Karanjkar
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16769v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have emerged as a powerful, differentiable class of learning models for graph-structured systems. Their ability to generalize across topologies opens the prospect of a surrogate for combined structural and parametric optimi...
65. MultiLoReFT: Decoupling Shared and Modality-Specific Subspaces in Multimodal Learning via Low-Rank Representation Fine-Tuning ​
Author: Sana Tonekaboni, Viktoria Schuster, Caroline Uhler
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16789v1 Announce Type: new Abstract: Real-world perception and decision making are inherently multimodal, integrating complementary signals across modalities. However, training multimodal models faces two main obstacles. First, collecting large-scale, well-aligned paired multimodal datase...
66. Value-Monotonicity Matters: A Concordance Loss for Deep Survival Prediction ​
Author: Meixu Chen, Kai Wang, Jing Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, stat.ML
arXiv:2607.16802v1 Announce Type: new Abstract: Deep survival models are evaluated almost exclusively by the concordance index (C-index), yet they are commonly trained using likelihood objectives such as the Cox partial likelihood, discrete-time negative log-likelihood, and DeepHit likelihood. This ...
67. Interpretable Anomaly and Drift Detection with Gaussian Mixture Models ​
Author: Behnam Asadi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16811v1 Announce Type: new Abstract: We revisit Gaussian Mixture Models (GMMs) as a lightweight, interpretable tool for anomaly detection and, in particular, for detecting distributional drift in data streams. We make three practical choices explicit and evaluate them on seven public benc...
68. Honest Physical-Support Inference after Latent Dictionary Learning: Collision Singularities and Minimax Resolution ​
Author: Guan-Ju Peng
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.TH
arXiv:2607.16813v1 Announce Type: new Abstract: Sparse-support uncertainty is usually quantified by treating the dictionary as known, an assumption that can produce overconfident, label-dependent conclusions when the dictionary is learned from latent sparse mixtures. Near collisions of coherent atom...
69. First-Order Predictable but Pairwise Fragile: Local Task Adaptation in Trained Transformers ​
Author: Irina Piontkovskaia, Sergey Nikolenko
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16821v1 Announce Type: new Abstract: Task arithmetic, sequential fine-tuning, activation steering, and first-order random search all operate through relatively small perturbations around an already trained checkpoint, and they rely on different local approximations: individual perturbatio...
70. Beyond Memory Leaderboards: Evaluating Scientific Memory as Budgeted Context Restoration ​
Author: Maksim Sheverev, David Finkelstein, Sergey Nikolenko
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16848v1 Announce Type: new Abstract: Long-term memory is becoming a core component of LLM agents, but most memory benchmarks evaluate conversations or compact summaries, while research agents need to restore evidence from full scientific papers. We introduce two full-text scientific-memor...
71. Principled Direction-Free Intrinsic Motivation through Model-Free Epistemic Free-Energy Estimators ​
Author: Alireza Furutanpey, Schahram Dustdar
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16858v1 Announce Type: new Abstract: Across environments with mixed sources of uncertainty, unsupervised reinforcement learning requires intrinsic motivation that does not precommit to a particular direction of surprise. Surprise minimization is scoped by design to ``unstable'' environmen...
72. Bridging battery design and health assessment through virtual sensing and physics-informed learning ​
Author: Wendi Guo, S{\o}ren Byg Vilsen, Daniel Ioan Stroe, Yaqi Li, Yicun Huang, Ashima Verma, Daniel Brandell
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2607.16864v1 Announce Type: new Abstract: Supercharging of lithium-ion batteries (LiBs) requires robust health monitoring to ensure durability, safety, and user confidence, particularly for emerging vehicle-to-grid applications with bidirectional energy flows. Yet battery management remains la...
73. HyBDM: Multi-Scale Hybrid Experts for Time Series Forecasting with Bidirectional Dependency Modeling ​
Author: Wenqiang Ma, Chen Cheng, Xue Cheng, Jiarui Ye
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16882v1 Announce Type: new Abstract: Time series forecasting (TSF) is vital to many applications, yet existing models often struggle to capture the heterogeneous long-range global patterns and short-range local variations in multivariate time series. While some approaches partially model ...
74. Certified-Gap Dual-Price Policies for Real-Time Truckload Bid Acceptance with Relocating, Clock-Constrained Resources ​
Author: Aswin Chandrasekaran
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC
arXiv:2607.16891v1 Announce Type: new Abstract: A truckload carrier must accept or reject each load tender within seconds. The decision depends on fleet state, hours-of-service (HOS) clocks, and appointment windows. We model this as a weakly coupled dynamic program in which the resources relocate an...
75. TVGL-CFM:Generating and Forecasting Time-Varying Trajectories of Dynamic Networks with Conditional Flow Matching ​
Author: Om Roy, Yashar Moshfeghi, Keith Malcolm Smith
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16894v1 Announce Type: new Abstract: Many complex systems such as brain networks, financial markets, and gene-regulatory circuits are described not by a fixed graph but by one that changes over time. A standard way to summarise such structure at each instant is the sparse precision (inver...
76. When Can Safe Controllers Adapt? Information before Commitment ​
Author: Venkatesh Saligrama
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY
arXiv:2607.16895v1 Announce Type: new Abstract: Safe adaptive control is online adaptation under a safety guarantee on the learning trajectory itself. The controller may use any causal, history-dependent rule and act differently across environments as data arrive. Only its safety guarantee is unifor...
77. Enhancing Personalized Bladder Cancer Treatment Through Reinforcement Learning: A Recurrent Patient State Transition Decision Support Framework ​
Author: Divyansh Chawla, Anshu Garg, Isshaan Singh
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16916v1 Announce Type: new Abstract: Bladder cancer treatment requires personalized and adaptive decision-making, particularly for recurrent disease, where treatment effectiveness changes across successive clinical episodes. Conventional clinical decision support systems typically rely on...
78. Investigation of Polycystic Ovary Syndrome (PCOS) Diagnosis Using Machine Learning Approaches ​
Author: Al Zadid Sultan Bin Habib, Md Asif Bin Syed, Md. Ekramul Islam, Tanpia Tasnim
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV
arXiv:2607.16941v1 Announce Type: new Abstract: Polycystic Ovarian Syndrome (PCOS) is a widespread hormone problem for women of childbearing age. Women with PCOS may not ovulate; they might have high levels of androgens and have many small cysts on the ovaries. It can cause missed or irregular menst...
79. CADENCE: Closing the Reasoning Gap via Coverage-Adaptive On-Policy Distillation ​
Author: Satyam Kumar, Saurabh Jha
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16955v1 Announce Type: new Abstract: On-policy knowledge distillation transfers reasoning from large teachers to compact students, but existing approaches suffer three compounding failure modes: (i) cold-start collapse, where a fresh student assigns near-zero mass to teacher-preferred tok...
80. SurvCF(t): Counterfactual Explanations for Survival Analysis in Predictive Maintenance Multivariate Time Series Data ​
Author: Zara Karazian, Panagiotis Papapetrou, Sindri Magn'usson, Erik Frisk, Tony Lindgren
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16969v1 Announce Type: new Abstract: Predictive maintenance relies on accurate Remaining Useful Life estimation, often formulated using survival analysis over multivariate time-series data. While modern deep survival models achieve strong predictive performance, their black-box nature lim...
81. TurboVec: A Case Study in Cost-Efficient Private Retrieval for Enterprise RAG via Codebook-Oblivious Quantization ​
Author: Navnit Shukla, Kamal Pandey, Omsankar Tiwari
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR
arXiv:2607.16973v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems increasingly power enterprise LLM applications, yet the vector retrieval layer introduces two underexplored challenges: (1) trained codebook quantizers may expose corpus statistics during index construction,...
82. Periodic Bootstrap Thompson Sampling For Periodically Non-Stationary Bandit Problems ​
Author: Boning Shao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.16986v1 Announce Type: new Abstract: This paper introduces Periodic Bootstrap Thompson Sampling (PBTS), an innovative extension of the classic Thompson Sampling (TS) algorithm tailored for bandit problems with periodic non-stationarity. Conventional TS accumulates all past observations, l...
83. Counterfactual Shapley Credit Assignment ​
Author: Mingxuan Li, Kaizhan-Lee, Elias Bareinboim
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16999v1 Announce Type: new Abstract: The Credit Assignment Problem (CAP) is fundamental to developing efficient and explainable Reinforcement Learning (RL) agents. Existing frameworks, whether relying on temporal contiguity or hindsight-conditioned reward reweighting, frequently fail to a...
84. Scalable Causal Imitation Learning ​
Author: Eylam Tagor, Mingxuan Li, Elias Bareinboim
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17003v1 Announce Type: new Abstract: Imitation learning enables learning a policy in an unknown environment with a latent reward signal using expert demonstrations, but it struggles when the imitator's and expert's observations are mismatched and unobserved confounders are present in expe...
85. Regularize or Localize: When Training-Time KV-Cache Geometry Pays Under Quantization ​
Author: Libo Sun, Po-Wei Harn, Zewei Zhang, Peixiong He, Xiao Qin
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17019v1 Announce Type: new Abstract: We study whether \sigreg -- LeJEPA's anti-collapse objective -- can reshape representations during standard autoregressive language-model pretraining, and when the resulting geometry helps \kv-cache quantization. We train 110M-parameter models on 10B F...
86. Interpretable Machine Learning for Air Pollution and Respiratory Health Prediction: A Socioeconomic Subgroup Analysis ​
Author: Maede Azani Hassan Abadi, Shouyi Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17024v1 Announce Type: new Abstract: Air pollution and climate-related stressors are increasingly important concerns for respiratory health, especially in settings with unequal environmental exposure and healthcare capacity. This study evaluates an interpretable machine learning framework...
87. ChemFusion: A Multimodal Cross-Attention Network for Reaction Yield Prediction ​
Author: Qiwei Han, Chi Zhou
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph
arXiv:2607.17033v1 Announce Type: new Abstract: Forecasting the outcomes of transition-metal-catalyzed reactions is notoriously complex due to the interplay of diverse physical and chemical variables. A persistent computational bottleneck has been effectively merging broad electronic descriptors wit...
88. Apeliotes: A Diffusion-Based Modeling Framework for km-scale Multi-Level Atmospheric Fields ​
Author: Evangelia Rafaela Frastali, Achyut Paudel, Maryam Golbazi, Frank Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2607.17037v1 Announce Type: new Abstract: High-resolution atmospheric data are required to resolve mesoscale and localized meteorological structures, however such datasets remain limited in many regions of the world. Existing high-resolution weather products are typically produced through dyna...
89. Solver-Hard Is Not Model-Hard: A Hardness-Controlled Diagnostic for LLM Constraint Reasoning ​
Author: Lucky Verma
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.LO
arXiv:2607.17047v1 Announce Type: new Abstract: LLM constraint reasoners are often evaluated near the random-SAT phase transition, confounding density and solver hardness. We test instance-level transfer while near-matching clause density. At aligned size bins, with near-matched density and matched ...
90. What does a Bayes-filtered transformer believe? A predictive Monte Carlo approach ​
Author: Afiq Abdillah Effiezal Aswadi, Haotong Ma, Susan Wei
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17060v1 Announce Type: new Abstract: A Bayes-filtered transformer (BFT) is a transformer trained on sequences that are generated in two steps: first a latent task is drawn from a prior, then observations are drawn conditional on that task. Trained under autoregressive log loss, the BFT's ...
91. Persistent Sparse Autoencoders: Learning Feature Timescales in Language Models ​
Author: Haoyan Luo, Mateo Espinosa Zarlenga, Mateja Jamnik
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.17117v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) decompose language model activations into sparse features, but standard SAEs encode each token independently and do not expose information that persists across a sequence. We introduce Persistent Sparse Autoencoders (Persiste...
92. Robust Assamese Speech Recognition through Controlled Fine-Tuning of Whisper Models ​
Author: Ganapati Das, Dwipen Laskar, Hasin Afzal Ahmed, Sanjib Kr Kalita, Kshirod Sarmah, Hem Chandra Das, Manjula Kalita
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17164v1 Announce Type: new Abstract: Developing Automatic Speech Recognition (ASR) for morphologically rich, low-resource languages such as Assamese is challenging due to insufficient annotated speech data. The pretrained Whisper model performs poorly on Assamese speech recognition tasks....
93. Explaining and Tuning Transformer-based LLMs in Arithmetic Tasks with Human Strategies ​
Author: Luyu Qiu, Jianing Li, Hwanhee Kim, Xiaoyong Wei, Yueyuan Zheng, Janet Hsiao, Lei Chen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17166v1 Announce Type: new Abstract: Transformer-based large language models (LLMs) continue to achieve state-of-the-art performance across various natural language processing tasks. However, their subpar performance on seemingly elementary problems, such as basic arithmetic, raises conce...
94. DADIR: Density-Aware Data-level Imbalanced Regression Framework ​
Author: Shermin Shahbazi, Hossein Mohammadi, Mohsen Afsharchi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17178v1 Announce Type: new Abstract: Imbalanced learning addresses predictive modeling problems with underrepresented regions of the data distribution. Although widely studied in classification, imbalanced regression remains challenging because of continuous target variables and heterogen...
95. DynImmune-BERT: Dynamic Immune Repertoire Modeling with Neural ODE Driven Continuous Transformers ​
Author: Rong Fu, Yongtai Liu, Xiaowen Ma, Haoyu Zhao, Shuo Yin, Yiqing Lyu, Long Zhang, Wangyu Wu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17244v1 Announce Type: new Abstract: Longitudinal T cell receptor repertoires contain signals of clonal expansion, contraction, disappearance, and reappearance after immune perturbation. Static repertoire language models usually summarize a sample as a bag of sequences, so the sampling in...
96. Distilled Reinforcement Learning for LLM Post-training ​
Author: Chen Wang, Zhaochun Li, Jionghao Bai, Yining Zhang, Hexuan Deng, Ge Lan, Yue Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17247v1 Announce Type: new Abstract: Large language model (LLM) post-training is essential for improving reasoning, adaptation, and alignment. Existing methods mainly follow two paradigms: reinforcement learning (RL) and on-policy distillation (OPD). However, RL relies on coarse-grained o...
97. Node4All: Learning Node Representation Beyond Datasets ​
Author: Dooho Lee, Jaemin Yoo
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2607.17272v1 Announce Type: new Abstract: Node representation learning has advanced rapidly, yet most existing methods rely on per-dataset training and hyperparameter tuning. This dataset-specific optimization comes from the difficulty of designing reusable graph models that generalize across ...
98. AIGB-R1: Self-Evolving Generative Auto-Bidding via Hierarchical Planner-Executor Optimization ​
Author: Yuejia Dou, Hesong Wang, Xinyu Zhang, Tianyu Wang, Zhilin Zhang, Chuan Yu, Jian Xu, Bo Zheng, Qi Qi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17281v1 Announce Type: new Abstract: Auto-bidding plays an essential role in online advertising, automatically adjusting bids for advertisers to optimize their commercial goals. The emerging AI-Generated Bidding (AIGB) paradigm widely adopts generative modeling to optimize bidding strateg...
99. An Iterative Geometric Approach to Optimizing Separating Hyperplanes ​
Author: Akos Hajnal
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17282v1 Announce Type: new Abstract: Given a binary-labeled linearly separable dataset, and the objective is to compute the maximum-margin separating hyperplane, also known as the hard-margin Support Vector Machine (SVM) classifier. This paper investigates whether, if given an initial sep...
100. Lookahead Branching for Neural Network Verification ​
Author: Liam Davis, Duo Zhou, Huan Zhang, Guy Katz, Clark Barrett, Haoze Wu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.LO
arXiv:2607.17290v1 Announce Type: new Abstract: In this work, we investigate the effect of lookahead branching strategies in neural network verification. We present a general recipe to integrate lookahead into any branch-and-bound verifier and demonstrate how one of the current state-of-the-art bran...
101. DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments ​
Author: Jun Nie, Zhiqin Yang, Zhenheng Tang, Yonggang Zhang, Xiaowen Chu, Xinmei Tian, Bo Han
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IR
arXiv:2607.17291v1 Announce Type: new Abstract: Deep research agents increasingly operate over the open web, where relevant records coexist with redundant summaries, outdated reports, and misleading documents. Existing evaluations offer limited insight into whether agents preserve sound evidential s...
102. WAR: Workload-Aware Rollouts for Synchronous Agentic Reinforcement Learning ​
Author: Ryan Xu, Atlas Zhao, David Bao, Frank Du
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.OS
arXiv:2607.17299v1 Announce Type: new Abstract: Long-horizon rollout generation has become the dominant systems bottleneck in agentic reinforcement learning (RL). As agents interact with environments over many turns, trajectories rapidly grow to tens of thousands of tokens, making synchronous RL tra...
103. Rationalizing Boltzmann Rationality: An Axiomatic Characterization of Entropy-Regularized Policies ​
Author: Silviu Pitis
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, econ.TH
arXiv:2607.17316v1 Announce Type: new Abstract: The softmax policy $\pi(a \mid s) \propto \exp(\beta Q(s,a))$ is the default model of stochastic choice in reinforcement learning (RL). Various justifications based on robustness, exploration, and optimization have been offered in the RL literature, bu...
104. TAPAS: Throughput-adaptive Perception for Autonomous Systems ​
Author: Aman Vyas, Vasista Kodumagulla, Zain Taufique, Pasi Liljeberg, Anil Kanduri
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17317v1 Announce Type: new Abstract: Autonomous systems rely on a perception module to navigate through dynamic environments. In real-world scenarios, the perception module's throughput requirements vary at runtime due to changes in scene complexity. However, existing perception strategie...
105. Rethinking the Suitability of Reinforcement Learning Algorithms Under Practical Transfer Constraints ​
Author: Hany Hamed, Abhishek Naik, Colin Bellinger, A. Rupam Mahmood
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.RO
arXiv:2607.17326v1 Announce Type: new Abstract: Transfer-oriented reinforcement learning requires evaluating algorithms along dimensions that go beyond standard sample efficiency. We focus on two dimensions: practical efficiency, which asks whether conclusions about algorithm suitability change unde...
106. When Drift Detectors cry Wolf: False Alarm Rates in continuous ML Monitoring ​
Author: Raj Shekhar Singh
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17336v1 Announce Type: new Abstract: Drift detection is a core component of production machine learning monitoring systems, where detectors are used to compare incoming data with a reference distribution and trigger alerts when changes occur. However, these detectors are often evaluated i...
107. A multiverse-consensus pipeline for reproducible feature selection in untargeted LC-MS metabolomics ​
Author: Mohammed Saeed Al-Huraibi, Ihsan Yozgat, Ahmet Kaplan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, q-bio.GN
arXiv:2607.17345v1 Announce Type: new Abstract: Background: Untargeted LC-MS metabolomics requires a long chain of preprocessing decisions, each with several equally defensible options. Analysts typically commit to one pipeline and report the resulting feature shortlist. How strongly that shortlist ...
108. Chebyshev Manifold Adaptation ​
Author: Jiawen Li
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17377v2 Announce Type: new Abstract: The paper presents a new parameter-efficient adaptation method called ChebyMA (Chebyshev Manifold Adaptation). ChebyMA adopts weight matrices through a multi-surface superposition of Chebyshev polynomial bases evaluated on learnable coordinates and com...
109. CoEvoP&R: Co-Evolving Placement Objectives with Routing Feedback via Large Language Models ​
Author: Ruogu Chen, Weihua Xiao, Ramesh Karri, Jie Han
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2607.17398v1 Announce Type: new Abstract: Analytical placers rely on differentiable objective functions to guide placement, typically combining intermediate surrogate metrics such as half-perimeter wirelength (HPWL) and cell-density penalties. However, these placement-stage surrogates remain m...
110. CORAL: Learning Amyloid Fibril Ligand Docking with Cooperative Binding Rewards ​
Author: Yasheng Sun, Bohan Li, Youqi Tao, J"urgen Schmidhuber
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17412v1 Announce Type: new Abstract: A hallmark of neurodegenerative diseases such as Alzheimer's and Parkinson's is the aberrant aggregation of proteins into amyloid fibrils, and small molecules that selectively bind to these fibrils hold promise as diagnostics, imaging probes, and thera...
111. Grounded verification of chemical and materials reasoning: detection is the bottleneck ​
Author: Can Polat, Mustafa Kurban, Erchin Serpedin, Hasan Kurban
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, physics.chem-ph, physics.comp-ph, quant-ph
arXiv:2607.17417v1 Announce Type: new Abstract: Large language models confabulate chemical objects (molecular formulas, space groups, formation energies) in fluent reasoning traces, concentrated on long-tail entities where confidence is least trustworthy. Deterministic, database-grounded verificatio...
112. Kernelized Linear Attention: Breaking the Capacity Wall with Symmetric Cones ​
Author: Ayoub Ghriss, Sourav Chakraborty
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT, stat.ML
arXiv:2607.17419v1 Announce Type: new Abstract: Linear attention promises constant-time recurrent inference but degrades sharply on associative recall. We formulate attention recall as a spherical-packing problem and introduce Kernelized Linear Attention Activations (KATA), a framework whose feature...
113. Decoder-Preserving Sparse Autoencoders: Which Readouts Survive Sparse Compression? ​
Author: Aniket Deshpande
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17425v2 Announce Type: new Abstract: Sparse autoencoders (SAEs) compress model activations into sparse codes, but equal reconstruction error and sparsity can preserve different linearly decodable signals. We formalize this ambiguity as a matrix-valued distortion between optimal ridge-pred...
114. Abliteration Is Not a Scalpel: Off-Target Effects of Refusal Removal on Decision Disposition Across Model Families ​
Author: Aleksander Fafu{\l}a
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, q-fin.CP
arXiv:2607.17427v1 Announce Type: new Abstract: Abliteration - deleting a model's refusal direction from its weights - is the standard recipe behind popular "uncensored" open-weight models. We show the surgery is not clean. As a disposition probe we use 21,600 decisions under uncertainty - weekly up...
115. Calibrated Alzheimer's Conversion Risk in Mild Cognitive Impairment: Persistent Homology of Clinical Trajectories with Conformal Guarantees ​
Author: Navin Bondade
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.17442v1 Announce Type: new Abstract: Background. Predicting conversion from mild cognitive impairment (MCI) to Alzheimer's disease (AD) is central to trial enrichment and care planning, yet existing models provide no individual-level uncertainty estimates and rarely include transparent le...
116. Residual-Guided Multi-Resolution Refinement of Foundation Models: A Case Study in Drought Forecasting ​
Author: Wentao Gao, Jiuyong Li, Lin Liu, Thuc Duy Le, Jixue Liu, Yanchang Zhao, Yun Chen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17507v1 Announce Type: new Abstract: Regional climate prediction presents unique challenges for time series foundation models, which typically process temporal patterns through single-pass inference. Expert climatologists, in contrast, employ multi-scale temporal analysis and iterative re...
117. Retrieval-Augmented Interpretable Learning: Towards Task-Specific Zero-Shot Models in Healthcare ​
Author: Sazan Mahbub, Caleb Ellington, Zhiyuan Li, Yixin Yang, Souvik Kundu, Ben Lengerich, Eric P. Xing
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17508v1 Announce Type: new Abstract: We introduce Retrieval-Augmented Interpretable Learning (RAIL), a probabilistic meta-learning framework for zero-shot generation of task-specific interpretable models that synthesizes coefficient-space structure from natural-language task descriptions ...
118. Lightweight Wrappers for Adapting Time Series Foundation Models to Regional Drought Forecasting ​
Author: Wentao Gao, Jiuyong Li, Lin Liu, Thuc Duy Le, Jixue Liu, Yanchang Zhao, Yun Chen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17511v1 Announce Type: new Abstract: Large \emph{Time Series Foundation Models} (TSFMs) demonstrate strong zero-shot forecasting capabilities across diverse domains. However, their application to regional climate forecasting faces practical challenges: model weights are often proprietary,...
119. After the Euclidean Highway: Hyperbolic Expert AI as the Next Innovation ​
Author: Kwan Soo Shin, In Seok Kang, Munho Lee
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.17513v1 Announce Type: new Abstract: Expert domains are trees; the Euclidean transformer is not, diluting parent-child structure exponentially at depth. The hyperbolic turn left one question unasked: not how much of a network to curve, but where curvature may touch the gradient. Placement...
120. One-step lowest-variance selection in a Gaussian random-field model motivated by masked diffusion: Total correlation and a square root collision threshold ​
Author: Linjun Li
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT, math.PR
arXiv:2607.17522v1 Announce Type: new Abstract: Motivated by confidence-guided parallel unmasking in masked discrete diffusion, we study a single selection step in a stylized Gaussian random-field model. A locally dependent nonnegative score field represents position wise uncertainty, and the schedu...
121. FailureAtlas: A Taxonomy of Failure Modes in Multi-Provider LLM Serving Infrastructure ​
Author: Vishal Pandey, Gopal Singh
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.SE
arXiv:2607.17525v1 Announce Type: new Abstract: Multi-provider LLM gateways reverse proxies that route, load-balance, and rate-limit requests across foundation-model APIs have become critical production infrastructure. Yet the failure modes specific to this architectural layer remain undocumented, s...
122. Program Synthesis for Simulation-Based Inference: Joint Model Selection and Parameter Estimation ​
Author: Siddharth Mishra-Sharma
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2607.17540v1 Announce Type: new Abstract: Neural simulation-based inference enables parameter estimation for complex models, but typically requires the user to specify a simulator encoding a fixed model structure. We present a framework for joint model selection and parameter estimation that c...
123. Volatility-Aware Extreme Event Detection in High-Frequency Financial Markets ​
Author: Maorufa Zaman, Haris Md Sahed
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17555v1 Announce Type: new Abstract: Predicting extreme price movements in high-frequency financial markets is a challenging task due to non-stationarity, heavy-tailed return distributions, and severe class imbalance. In particular, rare but impactful events are often difficult to detect ...
124. CoCurve: Cross-Module Co-Pruning Curvature for Training-Free Structured LLM Pruning ​
Author: Zhiren Gong, Zihao Zeng, Zijie Wang, Tiantong Wang, Chau Yuen, Wei Yang Bryan Lim
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17568v1 Announce Type: new Abstract: Structured pruning compresses large language models (LLMs) by removing whole computational units, such as attention heads and feed-forward (FFN) channel groups. Most training-free methods, however, rank these units independently, implicitly treating th...
125. A Weisfeiler-Leman Characterization of Global-Attention Graph Transformers for Mixed-Integer Linear Programs ​
Author: Md Abrar Jahin, Craig A. Knoblock, Jay Pujara
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17570v1 Announce Type: new Abstract: Graph foundation models (GFMs) with global attention are increasingly used to represent mixed-integer linear programs (MILPs), aiming to capture structure beyond the locality of standard graph neural networks. We study their expressive power through gr...
126. AGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models ​
Author: Ruiyi Ding, Jie Li, He Kang, Ziyan Liu, Chengru Song, Yuan chen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.SY, eess.SY
arXiv:2607.17572v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is a powerful reinforcement learning algorithm for aligning generative models with human preferences. While successful in large language models~\cite{shao2024deepseekmathpushinglimitsmathematical}, its extensio...
127. ANNLib: A Development Framework for Efficient Approximate Nearest Neighbor Search ​
Author: Zheqi Shen, Jingbo Su, Zijin Wan, Yan Gu, Yihan Sun
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.IR
arXiv:2607.17582v1 Announce Type: new Abstract: Approximate Nearest Neighbor Search (ANNS) plays a pivotal role in modern deep learning pipelines. Recently, many ANNS systems have been proposed to either provide broad functionality or reach high performance. However, it is yet difficult to achieve b...
128. Concentration and Mean-Square Bounds for Contractive Stochastic Approximation: A Unified Elementary Approach ​
Author: Siddharth Chandak
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY, math.OC
arXiv:2607.17595v1 Announce Type: new Abstract: We establish mean-square and concentration bounds for stochastic approximation (SA) with arbitrary norm contractive mappings, under a multiplicative noise model where the noise may scale affinely with the norm of the iterates, and the iterates are pote...
129. Trustworthy Protein-Ligand Binding Affinity Prediction via Reliability-Aware Multi-Engine Fusion ​
Author: Yongchan Hong, Defu Cao, Wenjin Liu, Thomas Ku, Jordy Homing Lam, Emily Nguyen, Willie Neiswanger, Vsevolod Katritch, Yan Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17601v1 Announce Type: new Abstract: Accurate protein-ligand binding affinity prediction is central to computational drug discovery, yet modern docking engines frequently disagree without indicating which prediction to trust. Consensus scoring and ensemble methods improve mean accuracy bu...
130. Optimizing the Preconditioner: A Black-box Online-to-Nonconvex Conversion with Static Regret Minimization Oracles ​
Author: Haichen Hu, David Simchi-Levi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML
arXiv:2607.17607v1 Announce Type: new Abstract: We study whether stochastic nonconvex optimization can be reduced to ordinary static regret minimization in online convex optimization in a black-box manner. For smooth nonconvex objectives, our reduction maintains a predictable gradient tracker, while...
131. PoLoRA: A Preconditioned Orthogonalized LoRA Optimizer ​
Author: Nikhil Ghosh, Tetiana Parshakova, Robert M. Gower
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, math.OC
arXiv:2607.17620v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) makes finetuning large language models cheaper by adding to each weight matrix a trainable low-rank update parameterized as the product of two matrices. These matrices are usually trained with Adam, which treats them as a sin...
132. Can Transformers Really Do It All? On the Compatibility of Inductive Biases Across Tasks ​
Author: Damien Teney, Liangze Jiang, Hemanth Saratchandran, Simon Lucey
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17624v1 Announce Type: new Abstract: Transformers are remarkably versatile and their design is largely consistent across a variety of applications. But are they optimal for any given task or dataset? The answer may be key for pushing AI beyond merely scaling current designs. Method. We ...
133. TypiCore: A Hybrid Active Query Strategy for Class-Incremental Learning on Time Series ​
Author: Gabor Szucs, Samuel Jacsev, Marcell Nemeth, Davide Dalle Pezze, Gian Antonio Susto
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17632v1 Announce Type: new Abstract: Time series data play a pivotal role across numerous domains, including healthcare and manufacturing. In real-world environments, models must cope with distribution shifts over time, a challenge commonly addressed through Continual Learning (CL) techni...
134. Selectivity Matters: Source Node Influence Pruning for Unsupervised Graph Domain Adaptation ​
Author: Ridong Han, Yawen Shen, Zhongnian Li, Tongfeng Sun, Xinzheng Xu, Abdulmotaleb El Saddik
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17668v1 Announce Type: new Abstract: Unsupervised Graph Domain Adaptation (UGDA) aims to facilitate knowledge transfer from a labeled source graph to an unlabeled target graph by mitigating cross-domain distribution shifts. Existing methods primarily focus on node-level feature alignment ...
135. GeneSpeak-FP: Target and Compound Retrieval from Observed Cell-Level Perturbation Signatures ​
Author: Kseniia Vaniushkina, Jeongmin Lim, Jinyong Park
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17671v1 Announce Type: new Abstract: Large-scale single-cell perturbation atlases make it possible to ask an inverse question: given an observed transcriptional response, which annotated targets and compounds in a fixed library are most consistent with that response? We present \model, a ...
136. Beyond Objective Expressivity: Geometry Preservation in Multimodal Contrastive Learning ​
Author: Tillmann Rheude, Roland Eils, Benjamin Wild
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17673v1 Announce Type: new Abstract: Contrastive learning is increasingly moving toward settings with three or more modalities instead of image-text pairs. Yet, extending models from pairwise to higher-order multimodal alignment can introduce optimization and representation challenges. We...
137. Uncovering Latent Reasoning Strategies in Language Models ​
Author: Awni Altabaa, John Lafferty
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17674v1 Announce Type: new Abstract: A language model $p_\theta(y \mid x)$ trained on reasoning tasks learns to solve problems via multiple distinct strategies, yet these strategies are implicit and entangled within the model's response distribution. We study the problem of decomposing th...
138. Planning with Transformers: Chain of Computation and Structured Context Windows ​
Author: Ehsan Futuhi, Nathan R. Sturtevant
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17710v1 Announce Type: new Abstract: Large Language Models (LLMs) have had a remarkable impact across many areas of machine learning. However, recent studies have shown that they struggle to reliably solve planning problems. At the same time, theoretical results have shown that transforme...
139. MXSens: Sensitivity-Aware Mixed-Precision Quantization for Efficient LLM Inference ​
Author: Simla Burcu Harma, Danila Mishin, Zhengyuan Su, Ayan Chakraborty, Elizaveta Kostenok, Dongho Ha, Babak Falsafi, Martin Jaggi, Yunho Oh, Amir Yazdanbakhsh
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17733v1 Announce Type: new Abstract: 4-bit quantization enables efficient LLM inference, but suffers from significant accuracy degradation due to outliers. Prior work addresses this problem via data rotation or mixed-precision integer quantization, but often relies on software-managed sca...
140. Towards Reliable Zero-Shot Crowd Forecasting: Evaluating Time Series Foundation Models for Special Event Pedestrian Forecasting ​
Author: Ziteng Li, Yanan Xin, Tina Comes, Serge Hoogendoorn
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17758v1 Announce Type: new Abstract: Managing massive crowds during infrequent special events requires reliable real-time pedestrian-flow forecasting to ensure public safety and operational efficiency. However, supervised forecasting methods face limitations in these contexts due to scarc...
141. Generalize and Guide: Decomposing Rewards for Few-Shot Inverse Reinforcement Learning ​
Author: Ziyi Liu, Grace Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO
arXiv:2607.17760v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) provides a powerful framework for learning from demonstrations. However, real-world tasks often exhibit substantial natural variations (e.g., picking up mugs with varying shapes), making it impractical to collect de...
142. FIFA World Cup 2026 as a Contamination-Free Benchmark for LLM Forecasting Agents: Four Models, a Bookmaker, and 104 Matches ​
Author: Jiacheng Ding, Cong Guo, Jason Xu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DB
arXiv:2607.17765v1 Announce Type: new Abstract: We introduce WC2026-Agents, a benchmark and dataset for evaluating large language models (LLMs) as autonomous forecasting agents on real, future events. For every one of the 104 matches of the 2026 FIFA World Cup, four frontier models -- Claude Opus 4....
143. Feature Attribution-Based Explainability Analysis of Deep Learning Models in Predictive Process Monitoring ​
Author: Kseniya Sahatova, Rafael Seidi Oyamada, Xuefei Lu, Johannes De Smedt
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17783v1 Announce Type: new Abstract: Predictive process monitoring supports the optimization and control of operational business processes by forecasting the future state or outcome of ongoing cases. While deep neural networks have achieved strong performance for these tasks by modeling s...
144. The Concept of Representation in ML: Beyond Plato and Aristotle ​
Author: Gilad Landau, Aviv Keren
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17800v1 Announce Type: new Abstract: Representation is a central concept in modern machine learning, where it usually refers to internal encodings that support learning and generalization. As models scale and their capabilities become increasingly human-level, this representational langua...
145. Phasor Attention: Mean Root Square Normalization for Phase Manifold Preservation ​
Author: Sungwoo Goo, Hwi-yeol Yun, Sangkeun Jung
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17822v1 Announce Type: new Abstract: While Root Mean Square Normalization has become the de facto standard for accelerating modern sequence models, its reliance on the quadratic accumulation of independent scalars ($\sum x^2$) inherently triggers outlier-induced numerical instability, gra...
146. Theoretical Foundations of $\max$@$k$ Reinforcement Learning ​
Author: Riccardo Poiani, Martino Bernasconi, Andrea Celli
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17823v1 Announce Type: new Abstract: Reinforcement Learning is a cornerstone technique for modern large reasoning models. Usually, for difficult tasks such as code generation and theorem proving, the agent is evaluated by generating $K$ responses rather than sampling a single response, an...
147. Mobius Learning: Cyclic Depth Folding in Transformers ​
Author: Tongtian Zhu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.DC
arXiv:2607.17843v1 Announce Type: new Abstract: Transformer-based language models organize computation along an ordered depth axis, where shallow and deep blocks often develop distinct representational roles. We challenge the conventional view that these roles must remain tied to a block's position ...
148. Distributional Soft Bellman Operator under the Cram\'er Geometry ​
Author: Keru Wang, Yixin Deng, Yao Lyu, Stephen Redmond, Shengbo Eben Li
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17897v1 Announce Type: new Abstract: Distributional soft policy iteration (DSPI) provides an important framework for combining distributional reinforcement learning (DRL) with maximum-entropy control, in which the policy evaluation step is governed by a distributional soft Bellman operato...
149. The Art of Not Forgetting ​
Author: Ashmith Atmuri, Akshay Kumar, Yashaswini Rao Bhogarajula
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17944v1 Announce Type: new Abstract: We introduce CMP (Cognitive Memory Primitive), an architecture that represents inputs as sparse relational codes, stores them in a two-tier competitive memory, and learns entirely through local, gradient-free updates, with no backpropagation anywhere i...
150. A Geometric Perspective on Stabilizing Value Conflict Resolution ​
Author: Saket Reddy, Andy Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17946v1 Announce Type: new Abstract: Large Language Models (LLMs) often struggle to navigate value conflicts when trained with the compressed scalar rewards of Reinforcement Learning from Human Feedback (RLHF). To address this challenge, we investigate how chain-of-thought (CoT) reasoning...
151. Topological Signatures of Context-Level Reliability in TabPFN ​
Author: James Hu, Mahdi Ghelichi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML
arXiv:2607.17962v1 Announce Type: new Abstract: TabPFN is a transformer-based foundation model for tabular prediction that performs inference without task-specific training by conditioning on a support set and query inputs. Despite its strong empirical performance, its internal behavior on structura...
152. DiFA: Inference-Time Forward-Process Alignment for Diffusion Models ​
Author: Shigui Li, Delu Zeng
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17972v1 Announce Type: new Abstract: The prevailing inference framework for diffusion models formulates generation fundamentally as a problem of numerical integration. This perspective casts the model as an exact estimator, neglecting the inherent statistical uncertainty of the denoising ...
153. Harness Engineering for LLM-Driven GPU Kernel Generation ​
Author: Yue Shui, Chenyu Ma, Hangfei Xu, Shengzhao Wen, Yanpeng Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.17979v1 Announce Type: new Abstract: Large language models (LLMs) can assist GPU kernel generation, but their practical effectiveness depends on whether generated code can be reliably constrained, validated, profiled, and selected. This paper presents a harness-centered system for LLM-dri...
154. Information-Based Exploration via Random Features for Reinforcement Learning ​
Author: Waris Radji, Odalric-Ambrym Maillard
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.17981v1 Announce Type: new Abstract: Representation learning has enabled classical exploration strategies to be extended to deep Reinforcement Learning (RL), but often makes algorithms more complex and theoretical guarantees harder to establish. We introduce Random Feature Information Gai...
155. fSRD: Fuzzy Spectral Region Decomposition -- Automated Multi Operator Koopman Representations via an Adaptive Spectral Learning Architecture ​
Author: Charles Bokor, Mark Cary, Denise Morrey, Fabrizio Bonatesta
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, math.DS
arXiv:2607.17990v1 Announce Type: new Abstract: Highly nonlinear chaotic dynamical systems remain difficult to model due to fundamental trade-offs between complexity, expressivity, and data efficiency. Modern machine learning methods achieve strong predictive performance but often rely on a-priori s...
156. MADA-RL: Multi-Agent Debate-Aware Reinforcement Learning for Parameter-Efficient Reasoning in Compact Models ​
Author: Martino M. L. Pulici, Cuong Xuan Chu, Evgeny Kharlamov, Zifeng Ding, Volker Tresp, Yunpu Ma
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.MA
arXiv:2607.18006v1 Announce Type: new Abstract: Large language models achieve strong reasoning performance, but often at prohibitive training cost - a challenge that is especially acute for compact models ($\leq 4 , \mathrm{B}$ parameters) trained under limited budgets. We introduce MADA-RL, a post...
157. FlashPDE: A Drop-in Fused Triton Operator Library for Neural PDE Solvers ​
Author: Peiyu Zang, Bosen Xie, Ruoxiang Xu, Yongqiang Cai
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.MS
arXiv:2607.18020v1 Announce Type: new Abstract: Physics-Informed Neural Networks (PINNs) solve PDEs by incorporating physical constraints into neural-network training, but large-scale problems are limited by automatic-differentiation memory overhead and inefficient execution of grid-based PDE operat...
158. L1 Augmented Attention as an Improved Vector Similarity Metric ​
Author: Kurt Godden
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.18027v1 Announce Type: new Abstract: Scaled dot product attention conflates directional alignment and vector magnitude, limiting its effectiveness as a similarity metric in Transformer models. We introduce L1 augmented attention, a simple and computationally parallelizable modification th...
159. Adaptive Mamba Neural Operators ​
Author: Zeyuan Song, Zheyu Jiang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.AP, math.NA
arXiv:2607.18043v1 Announce Type: new Abstract: Accurately solving partial differential equations (PDEs) on arbitrary geometries and a variety of meshes is an important task in science and engineering applications. In this paper, we propose Adaptive Mamba Neural Operators (AMO), which integrates rep...
160. SEE: Structure-aware Exploring \& Exploiting for Long-horizon GUI Agent Trajectory Synthesis ​
Author: Zhuohang Fan, Beichen Zhang, Yuanfa Li, Changqiao Wu, Wei Liu, Jian Luan, Weigang Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18046v1 Announce Type: new Abstract: Graphical User Interface (GUI) agents powered by vision-language models hold promise for automating real-world mobile tasks. However, progress is limited by the lack of high-coverage, long-horizon interaction trajectories collected from element-rich an...
161. SGN: A Similarity-based Generative Network for Data Generation under Distribution Shift ​
Author: Jiaqi Zhu, Xincheng Chen, Yuncheng Wu, Zhaojing Luo, Beng Chin Ooi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18072v1 Announce Type: new Abstract: Generative models trained on a source domain often produce samples that are poorly aligned with shifted target domains, limiting their effectiveness for target-domain data augmentation. Although target-specific adaptation can reduce this mismatch, it t...
162. Sobek: Streaming Equivariant Tensor Product Convolutions ​
Author: Vladimir Choro\v{s}ajev, C'edric B'eny
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18074v1 Announce Type: new Abstract: Equivariant graph neural networks repeatedly apply edge-conditioned tensor-product convolutions over graph edges. Conventional implementations materialize edge-specific weights, messages, and adjoints, causing tensor-product workspace and memory traffi...
163. Generalised Bellman recurrence and three dualities in sequential decision-making ​
Author: Fernando E. Rosas, David Hyland, Daniel Polani
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18077v1 Announce Type: new Abstract: What gives the Bellman equation its form? We show that the recursive properties of optimal value functions follow from three conditions: that the dynamics decomposes through sufficient statistics, that the return decomposes recursively, and that the ag...
164. SelectInfer: Selective Neuron Loading and Computation for On-Device LLMs ​
Author: Huzaifa Shaaban Kabakibo, Eric Schniedermeyer, Artem Burchanow, Lin Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18081v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across a range of Natural Language Processing (NLP) tasks, but their high computational and memory demands pose significant challenges for deployment on resource-constrained edge de...
165. Enhancing Rubric-based RL via Self-Distillation ​
Author: Mingxuan Xia, Yuhang Yang, Chao Ye, Shuai Zhu, Shenzhi Yang, Guangcheng Zhu, Yuhang Zhang, Cheng Peng, Haobo Wang, Siqing Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18082v2 Announce Type: new Abstract: Rubric-based RL has recently shown promise in improving LLMs on open-ended tasks. A widely recognized limitation of rubric-based RL is limited exploration: criteria that no rollout manages to satisfy (Unexplored Criteria, UC) receive no optimization si...
166. The Label Complexity of Class-Conditional Coverage under Distribution Shift ​
Author: Weijia Han, Lisha Qu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.18088v1 Announce Type: new Abstract: Standard evaluation of many recognition systems contains distribution shift by construction, since benchmarks place disjoint conditions in the training and test splits. Under such a shift, split conformal prediction keeps marginal coverage near the nom...
167. Empowering On-Device Model Adaptation with an Edge AI Inference Accelerator ​
Author: Mateusz Piechocki, Alessandro Capotondi, Marek Kraft
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AR, cs.CV
arXiv:2607.18101v1 Announce Type: new Abstract: On-device model adaptation is essential to enable lifelong personalization on resource-constrained hardware, but compute, power, and memory limitations of such devices make end-to-end backpropagation impractical for modern deep neural networks. This wo...
168. LLM-as-a-Coach: Experiential Learning for Non-Verifiable Tasks ​
Author: Tianzhu Ye, Li Dong, Guanheng Chen, He Zhu, Xun Wu, Shaohan Huang, Furu Wei
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2607.18110v1 Announce Type: new Abstract: Reinforcement learning (RL) on open-ended tasks compresses an LLM's rubric-based evaluation into a scalar reward, discarding rich textual feedback and conflating responses with distinct quality profiles. We propose Experiential Learning (EL), which rep...
169. Manifold-Constrained Hyper-Connections for Parameter-Efficient Finetuning ​
Author: Valentijn Oldenburg, Floris de Kam, Bente Zuijdam, Lieve Eberson, Nicky van Zutphen, Stef de Wildt, Ivo Verhoeven
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18130v1 Announce Type: new Abstract: Most parameter-efficient finetuning (PEFT) methods adapt weights or activations, thus leaving one of the key Transformer components unchanged: residual connections. This paper investigates Manifold-Constrained Hyper-Connections (mHC), a generalisation ...
170. Do Language Models Dream of Binding Molecules? Benchmarking LLMs under Spatial Constraints ​
Author: Thomas MacDougall, Maksim Kuznetsov, Roman Schutski, Rim Shayakhmetov, Maxim Malkov, Vladimir Aladinskiy, Alex Aliper, Alex Zhavoronkov
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2607.18144v1 Announce Type: new Abstract: Structure-based drug design (SBDD) leverages the 3D structure of protein targets, often complemented by other spatial constraints, to generate candidate binding molecules. While diffusion models have dominated as a leading paradigm for high-quality 3D ...
171. Totally Positive Matrices and the Highest-Order Coefficients of the Characteristic Polynomial ​
Author: Tiago Closs, Leandro Farina
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, math.RA
arXiv:2607.18148v1 Announce Type: new Abstract: We investigate the extent to which totally positive matrices can be distinguished through the highest-order coefficients of their characteristic polynomials. To identify the most informative coefficients, we also employed neural-network classifiers tog...
172. Differentiable Logic Gate Networks for Low-Latency EEG Classification on Edge Devices ​
Author: Shyamal Y. Dharia, Stephen D. Smith, Camilo E. Valderrama
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18149v1 Announce Type: new Abstract: Real-time EEG classification on edge devices is bottlenecked by the floating-point arithmetic of conventional neural networks. We investigated Differentiable Logic Gate Networks (Diff-Logic) as a hardware-native alternative that compiles models into pu...
173. The Calibration Channel Determines the Bayes-Error Proxy: An Exact Law for Temperature-Induced Distortion ​
Author: Shreyas Pradeepkumar Khandale
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18162v1 Announce Type: new Abstract: The soft-label Bayes-error estimator beta(z) = E[min(z, 1-z)] of Ishida et al. estimates the irreducible error of a binary task directly from probability-valued labels. Recent work by Ushio et al. showed that this estimator is fragile when the probabil...
174. OR Else: A Differentiable Trust Region for Policy Optimization ​
Author: Chinmay Rane, Kanishka Tyagi, Michael Manry
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.18163v1 Announce Type: new Abstract: PPO and the GRPO baseline studied here use clipped surrogate objectives whose favorable-direction saturation introduces an abrupt change in the scalar objective's derivative. We ask whether Output Reset (OR), a smooth one-sided saturation rule, offers ...
175. A Continual Validation, Updating, and Decision-Making Framework for Self-Adaptive Digital Twins via Robust Model Predictive Control: A Case Study in Additive Manufacturing ​
Author: Yi-Ping Chen, Ying-Kuan Tsai, Vispi Karkaria, Seul Lee, Daniel Apley, Wei Chen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.ST, stat.TH
arXiv:2607.18164v1 Announce Type: new Abstract: Digital Twins rely on surrogate models to mirror physical systems in real time, yet these models can degrade as operating conditions evolve, a phenomenon known as concept drift. Maintaining surrogate fidelity under drift, particularly when models must ...
176. FlashRT: Agent Harness for Guiding Agents to Deploy Real-Time Multimodal Applications ​
Author: Krish Agarwal, Zhuoming Chen, Yanyuan Qin, Zhenyu Gu, Atri Rudra, Beidi Chen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.18171v1 Announce Type: new Abstract: Real-time multimodal applications, including voice agents and interactive video generation, compose heterogeneous models into pipelines whose efficient deployment requires application-specific decisions about placement, streaming, and intra-model paral...
177. Three-Body Scattering for Generative Modeling ​
Author: Peng Sun, Zhenglin Cheng, Deyuan Liu, Jun Xie, Xinyi Shang, Tao Lin
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2607.18198v1 Announce Type: new Abstract: Modern generative models typically rely on an adversarial critic, a prescribed noise-to-data path, or an autoregressive factorization. Instead, we show that a proper distributional energy can induce sample-level motion and provide direct regression sup...
178. Causal Discovery on Irregular Time Series ​
Author: Martim Penim, Ricardo Ribeiro Pereira, Jacopo Bono, Hugo Ferreira, M'ario A. T. Figueiredo, Pedro Bizarro
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ME
arXiv:2607.18226v1 Announce Type: new Abstract: Causal discovery methods have shown strong performance in temporal systems, but they typically rely on regular and discrete lag structures, limiting their applicability to regularly sampled data. However, many real-world tasks require dealing with irre...
179. A Survey on GNN-based Link Prediction: Techniques, Applications, and Challenges ​
Author: Chengcheng Sun, Yajie Song, Cheng Zhai, Jiayun Tian, Jia Yang, Xiaobin Rui, Jian Zhang, Zhixiao Wang, Philip S. Yu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SI
arXiv:2607.16198v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) have emerged as the leading paradigm for link prediction, enabling the inference of missing connections and the anticipation of potential future links. However, existing reviews lack systematic exploration specifically ta...
180. Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Agentic RL ​
Author: Darshan Deshpande
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.16204v1 Announce Type: cross Abstract: Recent growth in reinforcement learning (RL) has surfaced a need for diverse, specialized training environments. Hand-curated environments with fixed task and reward difficulties become ineffective signals as model performance improves, and sparse re...
181. Comparing Spectrogram Front-Ends for Abnormal Heart-Sound Detection with a Convolutional Neural Network ​
Author: Abhinav Pala, Dhanush Pala
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG, cs.SD, eess.AS
arXiv:2607.16220v1 Announce Type: cross Abstract: Heart disease kills a lot of people, and one cheap way to catch it early is by listening to heart sounds with a stethoscope, or better yet, just recording them and running them through a model. This project is a binary classification task: take a sho...
182. Overcoming the BCI Calibration Bottleneck: A Clinically-Grounded Architecture using Riemannian Alignment and Stochastic Weight Averaging ​
Author: Immanuvel Prathap Sagayaraju
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SP, cs.LG, q-bio.NC
arXiv:2607.16225v1 Announce Type: cross Abstract: Brain-Computer Interfaces (BCIs) face a severe calibration bottleneck due to cross-subject spatial covariance shifts and physiological artifacts. To enable zero-calibration BCI, a deep learning pipeline was engineered combining Per-Session Independen...
183. FinBench: Time-Gated Calibration and Uncertainty Benchmarking for Agentic Financial Forecasting ​
Author: Rishab Ghosh, Vinay Devarakonda
Published: 7/21/2026, 4:00:00 AM
Categories: stat.AP, cs.LG, q-fin.CP, q-fin.PM
arXiv:2607.16229v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as components of agentic systems that observe, plan, and act. In finance, even "assistive" systems become decision-relevant once their outputs are used to size trades or allocate risk. A key failure ...
184. Composable Verification Pipelines for Multi-Agent Systems ​
Author: Julian Alfredo Mendez, Andreas Br"annstr"om
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LO, cs.AI, cs.LG, cs.MA, cs.PL
arXiv:2607.16266v1 Announce Type: cross Abstract: Existing approaches for reasoning about action and change provide expressive semantics for modeling dynamic systems, in most cases built on top of logic programming systems. We introduce a modular framework for transition and trajectory verification ...
185. A Novel Hybrid Quantum Reservoir Computing (nHQRC) for Phase Transition Detection in Non-Equilibrium Dynamical Systems ​
Author: Manoj B. Bhatkar, Prashant M. Yawalkar
Published: 7/21/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, q-fin.CP, q-fin.ST
arXiv:2607.16281v1 Announce Type: cross Abstract: The analysis of highly non-linear stochastic data within non-equilibrium dynamical systems requires computational frameworks capable of detecting latent phase transitions before systemic structural breakdowns occur. Traditional Variational Quantum Al...
186. Depth Estimators Are Implicit Neural Fields for 3D Scene Geometry Inpainting and Reconstruction ​
Author: Yingzhao Jian, Zihao Lin, Hehe Fan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.16286v1 Announce Type: cross Abstract: The 3D geometry of real-world scene data is often incomplete. Mainstream methods use depth estimators to inpaint missing structure. However, their prediction results can be inconsistent with observed geometry, or unreliable on out-of-distribution dat...
187. It Depends on the Dataset: When a Brain-Encoding Model's Predicted Responses Beat Their Visual Backbone for Video Memorability ​
Author: Carson Rodrigues
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.16292v2 Announce Type: cross Abstract: Brain-encoding foundation models predict fMRI responses to video, audio, and text well enough to win the Algonauts 2025 challenge. We ask whether their predicted responses, obtained with no scanner, are a useful feature lens for a downstream human-be...
188. SLT: Robust Quantum Neural Networks for Noisy-Label Medical Image Classification via Supermartingale-based Label Transition ​
Author: Jun Zhuang, Mohammad Al Hasan, Yiyu Shi, Chaowen Guan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.16293v1 Announce Type: cross Abstract: Noisy-label learning in small-scale medical image classification is challenging and hinders the superiority of deep neural networks. Recent studies suggest that quantum neural networks (QNNs) have shown potential in limited-data regimes, yet their us...
189. Efficient EEG Seizure Detection Using INT8 Quantization, Channel Pruning, and Spiking Neural Networks ​
Author: Kartikey Ahlawat
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG
arXiv:2607.16296v1 Announce Type: cross Abstract: Continuous EEG monitoring for epilepsy is constrained by the limited power and memory budgets of wearable and implantable devices. Deep neural networks can detect seizures with high accuracy, but their computational cost and model size make them diff...
190. On Hardware-Aware Design and Optimization of Edge Intelligence ​
Author: Shuo Huai, Hao Kong, Xiangzhong Luo, Di Liu, Ravi Subramaniam, Christian Makaya, Qian Lin, Weichen Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2607.16297v1 Announce Type: cross Abstract: Edge intelligence systems, the intersection of edge computing and artificial intelligence (AI), are pushing the frontier of AI applications. However, the complexity of deep learning models and heterogeneity of edge devices make the design of edge int...
191. Depth-Regularized JEPA World Models Learn More Transferable Representations from Real Outdoor Robot Data ​
Author: Usman M. Khan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO
arXiv:2607.16314v1 Announce Type: cross Abstract: World models, especially based on JEPA architectures, have been shown to learn robust dynamics of various environments. However, learning from visually complex real-world data remains a challenge, especially in unpredictable outdoor environments. We ...
192. Monte Carlo Dropout Uncertainty and Entropy-Thresholded Selective Prediction for Architecture-Agnostic Brain Tumor MRI Triage ​
Author: Medhansh Sharma
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.16317v1 Announce Type: cross Abstract: Deep networks now subtype brain tumors on MRI about as well as specialist readers, yet accuracy is not what keeps them out of the clinic. What matters at the point of care is whether a model's confidence can be trusted to flag the cases it is likely ...
193. ECG-LLM: Foundation Model for ECG-Based Cardiac Reasoning ​
Author: Alexander Selivanov, Friederike Jungmann, Jan Kehrer, Karl-Ludwig Laugwitz, Eimo Martens, Daniel Rueckert
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SP, cs.CV, cs.LG, eess.IV
arXiv:2607.16323v1 Announce Type: cross Abstract: Electrocardiography (ECG) is an inexpensive, standard-of-care test for cardiac symptoms, but front-line triage often lacks immediate access to definitive imaging such as echocardiography (ECHO) or cardiac magnetic resonance (CMR). Furthermore, most e...
194. SGMCE: Segment-Grounded Morphological Concept Explanation for Malaria Parasite Species Identification in Thick Blood Smears ​
Author: Ahmed Tahiru Issah, Charles B. Delahunt, Carine Mukamakuza
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.16324v1 Announce Type: cross Abstract: Malaria diagnosis in endemic regions depends on species-level identification of Plasmodium parasites in thick blood smears, but deep learning detectors classify detections without providing morphological evidence for their predictions, limiting the a...
195. RegionFM: Interpretable Region-Based Brain MRI Classification Using Foundation Model Embeddings ​
Author: Wei Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.16325v1 Announce Type: cross Abstract: Foundation models provide powerful representations for brain MRI analysis, but their predictions remain difficult to interpret in anatomically meaningful terms. Clinical assessment of brain MRI is commonly organized around anatomically defined struct...
196. Lipschitz Continuity in Deep Learning: A Systematic Review of Theoretical Foundations, Estimation Methods, Regularization Approaches, and Certifiable Robustness ​
Author: R'ois'in Luo, James McDermott, Colm O'Riordan
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.16329v2 Announce Type: cross Abstract: Lipschitz continuity is a fundamental property of neural networks that characterizes their sensitivity to input perturbations. It plays a pivotal role in deep learning, governing \textbf{robustness}, \textbf{generalization} and \textbf{optimization d...
197. AEVAL: From Anecdotal to Deterministic Testing for Agentic Skill Workflows ​
Author: Tejas Singh Anand, Yuet Ying Christina Wang, Wanting Jiang, Steve Masson, Tian Zheng, Bingjie Zhou
Published: 7/21/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG, cs.PF
arXiv:2607.16345v2 Announce Type: cross Abstract: Modern agentic systems increasingly rely on skills: installable packages of natural language and code that teach an LLM agent to perform a domain task. As skill repositories grow, developers need automated quality signals on every change, yet evaluat...
198. Boundary-Seeking GAN-Augmented TabTransformer for Adversarially Robust Intrusion Detection ​
Author: Raihan Sultan Pasha Basuki, Aliyah Kurniasih
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.16348v1 Announce Type: cross Abstract: Machine learning-based intrusion detection systems (IDSs) often suffer from class imbalance and vulnerability to adversarial attacks, leading to degraded detection performance and reduced robustness. This study proposes a TabTransformer framework aug...
199. Joint-Embedding Predictive Architecture for Sensor-based Activity Recognition ​
Author: Mohd Halim Mohd Noor, Abdulrahman M. A. Baraka
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SP, cs.AI, cs.LG
arXiv:2607.16350v1 Announce Type: cross Abstract: Sensor-based human activity recognition (HAR) has achieved significant progressed in fully supervised learning settings. However, these supervised learning models rely on large amount of labeled data, which require labor-intensive collection and meti...
200. MTSSL: Meta-Thresholding Semi-Supervised Learning ​
Author: Shuyang Liu, Ziang Zeng, Ruiqiu Zheng, Jiazheng Wang, Zechen Liu, Wenxi Li, Zhou Yu
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.16363v1 Announce Type: cross Abstract: A large body of Semi-supervised Learning~(SSL) algorithms encounter the threshold $\tau$ to select pseudo-labels. The value of $\tau$ across different SSL algorithms can vary depending on the learning perspective, yet they may achieve similar perform...
201. AoA: Theorem Proving Agent over Abstract Syntax Tree of Redesigned Language ​
Author: Qiyuan Xu, Joshua Ong Jun Leang, Renxi Wang, Wenda Li, Haonan Li, Luke Ong, Conrad Watt
Published: 7/21/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG, cs.PL
arXiv:2607.16372v1 Announce Type: cross Abstract: Interactive theorem proving (ITP) underpins program verification and formalized mathematics, but its manual effort limits scalability. LLM-based proof agents promise to ease this effort, but their heavy token consumption and API cost remain a major o...
202. Disentangling Model and Human Data Uncertainty in Apparent Facial Age Estimation ​
Author: Andrei Foitos, Ivo Pascal de Jong, Matias Valdenegro-Toro
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.16378v1 Announce Type: cross Abstract: Estimating the apparent age of individuals from facial images is challenging due to the subjective nature of perception and the inherent variability of the data. We investigate the role of uncertainty estimation, attributing uncertainty jointly to a ...
203. Think, Plan, Paint: Layout-Aware Reasoning for Controllable Image Generation in Unified Models ​
Author: Junhao Liu, Jian-Wei Zhang, Tao Huang, Miles Yang, Zhao Zhong, Liefeng Bo
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.16409v1 Announce Type: cross Abstract: Unified Multimodal Large Language Models (MLLMs) offer a promising paradigm for unifying visual understanding and generation, yet they still struggle to follow complex spatial instructions and logical constraints in controllable image generation. To ...
204. Back to the museum: Investigation of the acceptance of Android Andrea with and without emotion simulation in a museum ​
Author: Marcel Heisler, Christian Becker-Asano
Published: 7/21/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.HC, cs.LG
arXiv:2607.16428v1 Announce Type: cross Abstract: For a second time, the android robot Andrea was set up at a public museum in Germany for six consecutive days to have conversations with visitors, fully autonomously. Building on previously gathered qualitative results, the robot was now capable of e...
205. Spectral-Morphological Attention U-Net: An Efficient Network for Active Wildfire Detection ​
Author: Yugong Zeng, Jonathan Wu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.16472v1 Announce Type: cross Abstract: Over the past decades, the frequency of global wildfires has been increasing steadily. Therefore, if the fire can be detected and precisely located at an early stage, the potential hazards caused by it can be minimized to the greatest extent. The mac...
206. Foresight Residual RL for Long-Horizon Robot Manipulation with Vision-Language-Action Models ​
Author: Yuhan Liu, Xinyu Zhang, Litao Liu, Abdeslam Boularias
Published: 7/21/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.16506v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies offer strong general-purpose manipulation priors, but often fail on tight-tolerance, contact-rich assembly due to long-horizon credit assignment and subtask coupling: a state that is geometrically successful for ...
207. From Modalities to Propositions: A Language-Centric Framework for Multimodal Intelligence ​
Author: Nadine Chang, Maying Shen, Shizhe Diao, Jialiang Wang, Jingde Chen, Thomas Breuel, Pavlo Molchanov, Rafid Mahmood, Jose M. Alvarez
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.CV, cs.LG, cs.MM
arXiv:2607.16560v1 Announce Type: cross Abstract: We propose a language representation for multimodal data in which any observation, whether image, video, or text, is expressed as a bag of atomic propositions, simple statements about the entities, actions, and relations in a scene. A global semantic...
208. Exact Network Surgery: Functional Invariance and Gradient Plasticity in Reactive Computational Graphs ​
Author: Abdallah Khemais (ISITCOM, University of Sousse)
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.PL
arXiv:2607.16568v1 Announce Type: cross Abstract: Function-preserving network growth techniques such as Net2Net and progressive stacking expand a model's capacity without destroying its learned function, but existing formulations either tolerate numerical perturbations or require a full rebuild of t...
209. Mapping Order in Semicrystalline Polymers using Machine Learning of Nanobeam Electron Diffraction ​
Author: Nicholas Marchese, Arthur R. C. McCray, Yael Tsarfati, Karen Bustillo, Adam Marks, Alberto Salleo, Colin Ophus
Published: 7/21/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG
arXiv:2607.16570v1 Announce Type: cross Abstract: Organic mixed ionic electronic conductors (OMIECs) are a promising class of polymer materials for applications spanning neuromorphic computation to energy efficient electronics and bioelectronics. Despite being highly tunable, the relationship betwee...
210. Harnessing disorder to decouple extension and shear in kirigami metamaterials ​
Author: Haomin Yu, Hanxun Jin, Mingxuan Bi, Mohammad Jafari, Feng Helen Long, Michael J Greenberg, Farid Alisafaei, Guy Genin
Published: 7/21/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.LG
arXiv:2607.16583v1 Announce Type: cross Abstract: Kirigami turns stiff sheets into compliant, shape-morphing structures, but its reliance on periodic cut patterns comes at a cost: correlated panel rotations couple extension to shear, so stretching one axis drives a parasitic shear that cannot be sup...
211. Backpropagation-Free Trunk Training via the Split Forward Gradients ​
Author: Tian Qin, Wei-Min Huang
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.16612v1 Announce Type: cross Abstract: Backpropagation makes training deep networks memory intensive because it must store intermediate activations. Forward-mode methods avoid this cost, but their gradient estimates become increasingly noisy as the number of trained parameters grows. We i...
212. De-floored Principal Component Regression: When Rank Selection Alone Is Insufficient for Prediction ​
Author: Peng Zhao
Published: 7/21/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.ME, stat.TH
arXiv:2607.16638v1 Announce Type: cross Abstract: Principal component regression (PCR) regularizes high-dimensional prediction by choosing a spectral cutoff, but rank selection cannot correct systematic inflation of the retained empirical eigenvalues. We study clean Gaussian random designs in which ...
213. How Do You Choose Your AI Component? An Interview Study of Secure AI Integration in Practice ​
Author: Mahzabin Tamanna, Elizabeth Lin, Sparsha Gowda, Laurie Williams, Dominik Wermke
Published: 7/21/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.CR, cs.IR, cs.LG
arXiv:2607.16660v1 Announce Type: cross Abstract: The increasing adoption of Large Language Models (LLMs) as AI components in modern software systems introduces distinct security risks to the software supply chain. While many considerations and safety mechanisms are in place for components of the tr...
214. OpenLanguageModel: Readable and Composable Small-Language-Model Pretraining for Education and Research ​
Author: Tavish Mankash, Vardhaman Kalloli, Keshava Prasad, Deepan Muthirayan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.SE
arXiv:2607.16669v1 Announce Type: cross Abstract: OpenLanguageModel (OLM) is an open-source PyTorch library for building and pretraining small language models while keeping their machinery visible. In OLM, model code reads like the architecture: components are ordinary modules, while Block, Residual...
215. Isotonic Conformal Prediction ​
Author: Daniel Bensimon, Sean Xiang Yu, Eric D. Kolaczyk, Archer Y. Yang
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2607.16675v1 Announce Type: cross Abstract: A point prediction that is well calibrated on average can still be systematically biased conditional on its own value, undermining its use in downstream decision-making. We consider two objectives for reliable uncertainty quantification: self-calibra...
216. The Value of Depth in Message Passing on Sparse Graphs: A Kesten-Stigum Dichotomy ​
Author: Aseem Raj Baranwal
Published: 7/21/2026, 4:00:00 AM
Categories: math.ST, cs.LG, math.PR, stat.ML, stat.TH
arXiv:2607.16676v1 Announce Type: cross Abstract: How deep does a graph neural network need to be on a sparse graph? We study its purest statistical form: node classification on the sparse contextual stochastic block model (CSBM) with average degree $\Delta=O(1)$, whose local weak limit is a broadca...
217. Semi-Supervised Conditional Diffusion via Label Augmentation ​
Author: Jin Su, Yuan Gao, Yong Zhou, Jian Huang
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.16685v1 Announce Type: cross Abstract: Conditional diffusion models have become a powerful and flexible framework for learning complex conditional distributions from labeled data. In practice, however, acquiring high-quality labels is costly and time-consuming, leaving large volumes of un...
218. A Causal Markov Condition for Value ​
Author: Olav Benjamin Vassend
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2607.16717v1 Announce Type: cross Abstract: This paper proposes a causal independence principle for value -- the value Causal Markov Condition (v-CMC) -- and develops the conceptual and mathematical foundations of a "causal value theory" linking causality and utility. After motivating a local ...
219. Semi-Supervised Conditional Generative Learning through Stochastic Interpolation and Sufficient Representations ​
Author: Changyu Liu, Yuling Jiao, Jian Huang
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.16725v1 Announce Type: cross Abstract: Conditional generative modeling remains a challenging problem in semi-supervised settings where labeled data is scarce but unlabeled samples are abundant. To effectively leverage structural information embedded within the unlabeled dataset and compen...
220. Dropout and Random Gradient Masking Are Asymptotically Equivalent in Large ResNets ​
Author: Javier Maass, L'ena"ic Chizat
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR
arXiv:2607.16761v1 Announce Type: cross Abstract: Dropout and Random Gradient Masking (RaM) are two training techniques used to improve performance in deep learning. Both techniques inject randomness into the training dynamics, but in significantly different ways: dropout applies random masks to the...
221. Demodulation of chaotic signals using convolutional neural network ​
Author: Mykola Kozlenko, Emrullah Demiral, Anton Yudhana
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2607.16788v1 Announce Type: cross Abstract: Chaotic modulation is an effective communication technique that exploits deterministic chaos to produce pseudo-random signals. A widely adopted approach involves modulation of the chaotic bifurcation parameter. This paper introduces a deep learning-b...
222. Diagnosing Correctness Probes under Self-Judgement Confounding ​
Author: Yi-Long Lu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.16799v1 Announce Type: cross Abstract: Hidden-state readouts can predict whether language-model outputs are correct, but objective correctness (OC) usually agrees with the model's own self-judgement (SJ), leaving the decoded signal semantically ambiguous. We construct conflict cases in wh...
223. Identity-Paired Progressive Depth Training: When Trainability Persists Beyond Expressibility ​
Author: Athanasios Hadjidimoulas, Tirthak Patel, Anastasios Kyrillidis
Published: 7/21/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, math.OC
arXiv:2607.16800v1 Announce Type: cross Abstract: Variational Quantum Algorithms (VQAs) are a leading paradigm for near-term quantum computing, yet their training suffers from sensitivity to circuit depth, initialization, and landscape pathologies such as barren plateaus. We study \emph{progressive ...
224. Explainable Lightweight Compact Deep Models for Speech Emotion Recognition ​
Author: Nelly Elsayed
Published: 7/21/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.ET, cs.LG
arXiv:2607.16803v1 Announce Type: cross Abstract: Speech Emotion Recognition (SER) is an important component in a wide range of human-centered applications, including healthcare, customer service, and human-omputer interaction. In medical and decision-support settings, there is increasing interest i...
225. Trace-Based On-Policy Distillation for Masked Diffusion Language Models ​
Author: Haolin Ren, Ziyang Huang, Chenhao Yuan, Jun Zhao, Kang Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.16872v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) are a promising alternative to autoregressive generation. However, reasoning-oriented post-training for dLLMs remains challenging. Supervised fine-tuning (SFT) for dLLMs requires dense but often off-policy mask...
226. Hierarchical Wireless Foundation Model for Multi-Task Optimization ​
Author: Yangjing Wang, Ouya Wang, Shenglong Zhou, Geoffrey Ye Li
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2607.16877v1 Announce Type: cross Abstract: The increasing complexity of next-generation wireless networks has driven the integration of artificial intelligence (AI) into wireless communications. However, most existing studies focus on developing task-specific deep learning techniques for sing...
227. Transferable Low-Rank Convolutional Bases for Onboarding Unseen Medical Imaging Modalities ​
Author: Ranat Das Prangon, Istiaque Ahmed, Shajid Hasan Naim, Waseem Mustak Zisan, Hossain Md Shakhawat
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.16888v1 Announce Type: cross Abstract: Deploying a medical imaging model that must later accommodate a modality it has never seen is a recurring practical problem: retraining the shared representation is expensive and destroys performance on the modalities already in service. We study thi...
228. A Method for Learning Value Systems in Generative AI ​
Author: Andr'es Holgado-S'anchez, Holger Billhardt, Sascha Ossowski
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CL, cs.LG
arXiv:2607.16903v1 Announce Type: cross Abstract: Value-aware AI systems require explicit computational representations of human values (groundings) and their aggregation into value systems in order to align their decisions with ours. As such representations are difficult to elicit, value learning s...
229. Deep Adaptive Bayesian Screening ​
Author: Jade Lejeune Herman, Arno Strouwen, Johan A. K. Suykens, Peter Goos
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.16927v1 Announce Type: cross Abstract: We introduce Deep Adaptive Bayesian Screening (DABS), a method for performing adaptive factorial screening in high-dimensional discrete design spaces. DABS learns a policy network offline to sequentially select informative experiments, amortizing Bay...
230. Optimizing Clinical Trial Protocols Using EHR-Derived Heterogeneous Treatment Effects ​
Author: Xiaodi Li, Munhuwan Lee, Pengyang Li, Xiaoke Liu, Jose K. James, Patricia A. Pellikka, Cui Tao, Nansu Zong
Published: 7/21/2026, 4:00:00 AM
Categories: stat.AP, cs.AI, cs.LG
arXiv:2607.16934v1 Announce Type: cross Abstract: Traditional randomized trials often obscure clinically meaningful heterogeneity in treatment response by focusing on average effects. Leveraging real-world data to emulate clinical trials and estimate heterogeneous treatment effects (HTEs) offers a p...
231. Pediatric Bone Age Prediction Using Deep Learning ​
Author: Al Zadid Sultan Bin Habib, Md. Ekramul Islam, Md Asif Bin Syed, Md Younus Ahamed, Tanpia Tasnim
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.16936v1 Announce Type: cross Abstract: Pediatric bone age prediction is a crucial task in clinical practice that can help diagnose endocrine disorders and provide insight into a child's growth and development. However, conventional bone age prediction methods are often labor-intensive and...
232. What Do They See? Interpreting Complex Road Scenarios Through the Eyes of Vision-Language-Action Models for Safe and Trustworthy Autonomous Vehicle Learning ​
Author: Kalpana Panda, Wesley Maia, Vinti Agarwal, Ross Greer
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.RO
arXiv:2607.16938v1 Announce Type: cross Abstract: End-to-end autonomous driving models are now able to navigate complex road scenarios, mapping raw sensor observations directly to observed paths for open-loop evaluation and often effective driving in closed-loop evaluation. Yet the internal logic of...
233. Tight Sample Bounds for Renyi and Min-Entropy Estimation ​
Author: Arman Adibi, Piotr Krysta
Published: 7/21/2026, 4:00:00 AM
Categories: cs.IT, cs.CC, cs.LG, math.IT, math.ST, stat.TH
arXiv:2607.16966v1 Announce Type: cross Abstract: Estimating entropy from samples is fundamental in information theory and property testing. Shannon entropy measures average uncertainty and can be estimated to constant additive accuracy over a $k$-symbol alphabet using $\Theta(k/\log k)$ samples. Mi...
234. Twisted Schr\"odinger Bridge Matching ​
Author: Maxence Noble, Marie Scheid, Yazid Janati, Eric Moulines, Alain Durmus
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.16987v1 Announce Type: cross Abstract: Over the past few years, diffusion-based Schr"odinger bridge models have been proposed to approximate optimal transport dynamics between two prescribed boundary distributions, with successful applications to generative modeling. More precisely, thes...
235. PriorProof: A Point-in-Time Measure of Technique Novelty for Formal Proofs ​
Author: Neel Somani
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO
arXiv:2607.16997v1 Announce Type: cross Abstract: Mathematicians distinguish proofs that explain, simplify, or introduce a nonstandard route, but these judgments are difficult to operationalize. We study a deliberately narrower construct: time-relative proof-route nonstandardness in formal mathemati...
236. Increasing Line Outage Localization Performance with Ensemble Classifiers ​
Author: Daniel Flores, Yuanrui Sang, Michael P. McGarry
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2607.17008v1 Announce Type: cross Abstract: In many cases, the outage of one transmission line in a system can be localized by monitoring the power flow of another line, and machine learning methods can be used to distinguish the cases under uncertainty. In this study, we examine the improveme...
237. WHALE: A Scalable Unified Model for Recommendation with Wukong-HSTU Architecture ​
Author: Renqin Cai, Dawei Sun, Yuanjun Yao, Zhiyong Wang, Velvin Fu, Maggie Zhuang, Yu Shi, Zhongnan Fang, Xuan Cao, Jing Qian, Rui Li
Published: 7/21/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG
arXiv:2607.17017v1 Announce Type: cross Abstract: As scalability becomes increasingly important in recommendation modeling, recent architectures have advanced the modeling of two broad sources of ranking signals along separate paths: non-sequence features, including user, item, context, and cross fe...
238. Robust Chance-Constrained Optimization using a Continuous Parameter Space Wasserstein-2 Ambiguity Set of Gaussian Mixtures ​
Author: Shibshankar Dey, Sanjay Mehrotra
Published: 7/21/2026, 4:00:00 AM
Categories: math.OC, cs.LG
arXiv:2607.17018v1 Announce Type: cross Abstract: We study distributionally robust linear chance-constrained problems in which uncertainty is modeled by a Gaussian mixture model (GMM). Finite-support distributionally robust (FDR) formulations, widely used in data-driven robust optimization, robustif...
239. Otap:Structure-Aware Optimal Transport for Evaluating Planning and Execution in Agent Trajectories ​
Author: Babak Barazandeh, Subhabrata Majumdar, George Michailidis
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.17082v1 Announce Type: cross Abstract: Large language model agents solve tasks by generating trajectories that interleave planning, tool calls, and intermediate results. Current evaluation metrics reduce such a trajectory to a binary success flag or compare it against a reference by exact...
240. ThRIve: Thermally Robust CNN Inference via Low-Rank Adaptation in Heterogeneous PIM Architectures ​
Author: Vibhanshu Sharma, Pratyush Dhingra, Janardhan Rao Doppa, Partha Pratim Pande
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AR, cs.ET, cs.LG
arXiv:2607.17091v1 Announce Type: cross Abstract: Processing-In-Memory (PIM) has emerged as a promising technology for accelerating machine learning (ML) workloads. Specifically, non-volatile memory-based PIM architectures have enabled effective ML acceleration due to their ability to perform energy...
241. A Multi-Model Hybrid Defense Approach Against White-box Adversarial Attacks in Computer Network Traffic ​
Author: Khushnaseeb Roshan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2607.17105v1 Announce Type: cross Abstract: It is crucial to safeguard computer networks from evolving network security threats and unknown cyberattacks. An essential tool for protecting computer networks against unknown cyber threats is Network Intrusion Detection System (NIDS). However, NIDS...
242. Teach it to stop, not just to click ​
Author: Barada Sahu (Cabal AI), Shivesh Pandey (Para AI)
Published: 7/21/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.HC, cs.LG
arXiv:2607.17136v1 Announce Type: cross Abstract: Agentic computer-use RL is reported in single runs, and those numbers mislead. Using verifier-guided repair of a 35B computer-use agent (CUA) across five oracle-graded environments, we show a repaired policy's success rate is dominated by upstream va...
243. The Geometry of Semantic Space: A Continuous Geometric Framework for the Transformer Architecture ​
Author: Zhihua Liang
Published: 7/21/2026, 4:00:00 AM
Categories: cond-mat.dis-nn, cs.CL, cs.LG
arXiv:2607.17146v1 Announce Type: cross Abstract: We present a continuous geometric framework that models the discrete algebraic operations of the Transformer architecture as an integro-differential equation (IDE) on a semantic fiber bundle $\calE = \calM \times \R^d$. Beginning from a single geomet...
244. OrderMoE: An expert similarity driven distributed edge MoE inference ​
Author: Xin Yuan, Ning Li, Quan Chen, Wenchao Xu, Athanasios V. Vasilakos, Song Guo, Haijun Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.NI, cs.DC, cs.LG
arXiv:2607.17154v1 Announce Type: cross Abstract: Although mixture-of-experts, MoE, models have been increasingly adopted to scale large language models with moderate computation cost, it remains challenging to deploy MoE inference over resource-constrained and bandwidth-limited edge infrastructures...
245. Non-Asymptotic Best Policy Identification Guarantees in Online Reinforcement Learning ​
Author: Joseph Lazzaro, Alessio Russo, Aldo Pacchiano
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.17201v1 Announce Type: cross Abstract: In this work we study the Best Policy Identification (BPI) problem in online, tabular Reinforcement Learning. This is an active sequential hypothesis testing problem in which the learner's objective is to identify an optimal policy in a Markov Decisi...
246. Rate-Distortion-Perception Theory: Redefining the Fundamental Limits of Information Representation ​
Author: Photios A. Stavrou, Giuseppe Serra, Marios Kountouris
Published: 7/21/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, cs.SY, eess.SY, math.IT
arXiv:2607.17232v1 Announce Type: cross Abstract: Classical rate-distortion (RD) theory has long established the fundamental limits of lossy compression by quantifying the minimum number of bits required to represent a source under a prescribed distortion constraint. However, widely used distortion ...
247. Lossless but Not Free: An Empirical Anatomy of Speculative Decoding on Consumer Hardware ​
Author: Param Chordiya
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.17283v1 Announce Type: cross Abstract: Single-stream autoregressive decoding of large language models is bound by memory bandwidth: each generated token requires one full forward pass through the target model, and successive passes cannot be parallelized. Speculative decoding restructures...
248. Interpreting Quantum Learning Models via Stochastic Processes ​
Author: Johannes Fankhauser, Lukas J. Fiderer, Hans J. Briegel
Published: 7/21/2026, 4:00:00 AM
Categories: quant-ph, cs.LG
arXiv:2607.17327v1 Announce Type: cross Abstract: Quantum machine learning models define probabilistic input--output maps through coherent quantum evolution and measurement. While such models can exhibit computational advantages, their internal functioning and decision making generally resists inter...
249. Self-Modifying Lean Proof Agents with Verifier-Grounded Benchmark Coevolution ​
Author: Yuqing Li, Zeguan Wu, Yu Gan, Junyu Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, quant-ph
arXiv:2607.17352v1 Announce Type: cross Abstract: Designing effective Lean proof agents is a central challenge in formal mathematical reasoning. Beyond building stronger provers, recent work emphasizes the workflow around Lean: how an agent decomposes proof obligations, uses tools and compiler feedb...
250. Stringological sequence prediction II: Right-to-left automaticity and related complexity measures ​
Author: Vanessa Kosoy
Published: 7/21/2026, 4:00:00 AM
Categories: cs.FL, cs.DS, cs.LG
arXiv:2607.17369v1 Announce Type: cross Abstract: In a previous paper, we began the study of sequence prediction algorithms adapted to stringological word complexity measures. One measure we considered was left-to-right (most-significant-digit-first) automaticity. Here, we show a statistically and c...
251. Taurus: Accelerating Out-of-Core Graph Neural Network Inference on Billion-Scale Graphs ​
Author: Pranjal Naman, Yogesh Simmhan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2607.17374v1 Announce Type: cross Abstract: Graph Neural Network (GNN) inference on billion-scale graphs is challenging due to the large memory footprint of features and embeddings and high disk I/O costs in out-of-core settings. Existing distributed GNN systems incur high communication times ...
252. Team DACTYL at PAN 2026: Bayesian Data Mixing and Empirical X-risk Minimization for AI-text Detection ​
Author: Shantanu Thorat
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.17382v1 Announce Type: cross Abstract: Existing research shows that AI-generated text detection classifiers achieve strong in-distribution (ID) performance but do not maintain the same performance on out-of-distribution (OOD) texts, suggesting overfitting to dataset-specific features. How...
253. Quantifying Diversity of Thought: A Predictive Law of Weighted LLM Ensemble Lift ​
Author: Junade Ali
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO, cs.MA
arXiv:2607.17384v2 Announce Type: cross Abstract: This paper provides an experimentally verified formal law for calculating the uplift that diversity of thought provides in Large Language Model (LLM) ensembles. From first principles, we derive an exact decomposition of LLM ensemble lift into rescue ...
254. Kernel Regression with Tensor Trains and Hadamard Overparameterization ​
Author: Duc Thien Nguyen, Konstantinos Slavakis, Eleftherios Kofidis, Dimitris Pados
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, eess.SP
arXiv:2607.17390v2 Announce Type: cross Abstract: Kernel regression with tensor trains and Hadamard overparameterization (KReTTaH) is introduced as a training-data-free, interpretable, and nonparametric framework for multi-way data imputation. The imputation problem is reformulated as regression in ...
255. Efficient Sequential Evaluation of Large Language Models ​
Author: Chia-Yu Hsu, Shubhanshu Shekhar
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2607.17409v1 Announce Type: cross Abstract: We study the problem of sequentially evaluating a new large language model (LLM) on a fixed question set using historical performance data from prior LLMs. Our goal is to construct a confidence sequence (CS) for the model's capability on this questio...
256. Feature-Guided Diffusion for Non-Differentiable Inverse Rendering ​
Author: Andrei-Timotei Ardelean, Michael Fischer, Tim Weyrich, Tom'a\v{s} Iser
Published: 7/21/2026, 4:00:00 AM
Categories: cs.GR, cs.LG
arXiv:2607.17411v1 Announce Type: cross Abstract: Inverse rendering is traditionally solved via differentiable renderers and gradient descent, which requires substantial problem-specific engineering and is prone to getting stuck in local minima due to ambiguities. Derivative-free approaches alleviat...
257. DA-MergeLoRA: Hypernetwork-Based LoRA Merging for Few-Shot Test-Time Domain Adaptation ​
Author: Siobhan Reid, Zhixiang Chi, Li Gu, Omid Reza Heidari, Ziqiang Wang, Yang Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.17467v1 Announce Type: cross Abstract: Few-shot Test-Time Domain Adaptation (FSTT-DA) seeks to adapt models to novel domains using only a handful of unlabeled target samples. This setting is more realistic than typical domain adaptation setups, which assume access to target data during so...
258. The Dimension of Nonterminating Resampling Computations ​
Author: Yunbei Xu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CC, cond-mat.stat-mech, cs.LG, math.CO, math.PR
arXiv:2607.17469v1 Announce Type: cross Abstract: A randomized algorithm may terminate almost surely even though exceptional random tapes make it run forever. This paper studies the survival tail, the Kolmogorov complexity of one such tape, and the Hausdorff dimension of all of them. For each $s>0$ ...
259. Panache: One-Pass Motif Discovery at Every Window Length ​
Author: Tej Sanibh Ranade
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.DB, cs.LG
arXiv:2607.17481v1 Announce Type: cross Abstract: Motif discovery, the search for recurring patterns within a time series, is a core primitive of exploratory data analysis. A pattern, however, is defined by its duration, which analysts rarely know in advance. To resolve this unknown duration, an int...
260. SALT: Salience-Aware Lexical Trie for Long-Context Compression ​
Author: Oteo Mamo, Hyunjin Yi, Joydhriti Choudhury, Shangqian Gao, Weikuan Yu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.PF, cs.AI, cs.LG
arXiv:2607.17486v1 Announce Type: cross Abstract: As large language models (LLMs) process increasingly longer prompts, computation and KV-cache memory costs have emerged as major bottlenecks in inference systems. Existing input-level prompt compression methods address this, but rank each sentence by...
261. Token-Level Off-Policy Learning for Faithful Generation Under Distribution Shift ​
Author: Zitong Huang, Gustavo Lucas Carvalho, Deqing Fu, Robin Jia
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.17524v1 Announce Type: cross Abstract: We propose Token-Level Off-Policy Labeling (TOPL), an off-policy training paradigm that reframes post-training as a token-level correctness prediction task. Our key intuition is that by training the model to distinguish good and bad tokens in a respo...
262. FlowSonic: Stable Zero-Shot Music Editing via High-Order Trajectory Integration ​
Author: Ali Boudaghi, Hadi Zare
Published: 7/21/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, cs.NA, eess.AS, eess.SP, math.NA
arXiv:2607.17526v1 Announce Type: cross Abstract: Zero-shot text-guided editing of real-world music recordings requires balancing semantic modification with faithful preservation of the original musical structure. Although recent diffusion transformers trained with rectified flow have achieved remar...
263. Can AI Agents Really Complete RTL-to-GDS? Lessons from Benchmarking Tool-Interactive EDA Workflows ​
Author: Jinyuan Deng, Zhengrui Chen, Xufeng Wei, Tianyu Xing, Chenyi Wen, Cheng Zhuo
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.AR, cs.LG
arXiv:2607.17528v2 Announce Type: cross Abstract: LLM-driven agent systems have emerged as a promising paradigm for electronic design automation (EDA), demonstrating strong potential for automating complex design workflows. However, existing evaluations primarily examine individual language models o...
264. The Curvature Shadow: An Apparent Failure of Maximum-Entropy Equilibrium Selection is a Removable Artifact ​
Author: Luis Leal
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.GT, cs.LG, cs.MA
arXiv:2607.17543v1 Announce Type: cross Abstract: In two-player zero-sum games whose Nash equilibria form a convex set, regularized solvers such as Regularized Nash Dynamics (R-NaD) empirically select the maximum-entropy member: the information projection (I-projection) of a uniform reference onto t...
265. COLIP-2: Olfaction-Vision-Language Embeddings ​
Author: Kordel Kade France
Published: 7/21/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.ET, cs.LG
arXiv:2607.17559v1 Announce Type: cross Abstract: The Contrastive Olfaction-Language-Image Pre-training 2 (COLIP-2) model is a multimodal embeddings space that places olfaction as a first-class citizen among vision and language. Molecular structure, gas-sensor readings, odor-descriptor language, and...
266. ZifaMem: Structured Memory for Persona, Preference, and Emotional Continuity in AI Companions ​
Author: Jingzhe Fang, Guozhi Xu, Yunfan Cui, Xiaochen Yang, Zhangyu Hua
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.17564v1 Announce Type: cross Abstract: AI companions are judged not only by single-turn fluency but by whether they sustain emotional continuity: remembering who the companion is, what the user prefers, and how the relationship has felt. We present ZifaMem, a structured memory system that...
267. Detection, Attribution, Narration: An End-to-End Pipeline for Explainable Money Mule Identification ​
Author: Yuge Zhang, Yuanxing Zhang, Yichao Jin, Khairul Amsyar Mohd Razis, Nicholas Qi An Choo, Kai Yin Anders Wong, Xinyan Tang, Kenneth Zhu Ke, Wee Keong Dennis Lee, Jingyuan Zhao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.17586v1 Announce Type: cross Abstract: Money mule accounts are critical facilitators of financial fraud, yet detecting them at scale remains challenging due to the heterogeneous nature of transactional and behavioural data. We present an end-to-end pipeline for customer-level mule detecti...
268. Semantic Color Naturalness Breaker: Preventing Illegitimate Colorization via Content-Aware Color Priors ​
Author: Yuki Nii, Futa Waseda, Ching-Chun Chang, Isao Echizen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.17610v1 Announce Type: cross Abstract: Automatic image colorization enables large-scale and low-cost reuse of grayscale media (e.g., manga panels and archival photographs), facilitating unauthorized reuse and redistribution. Once released online, grayscale content can be readily turned in...
269. Online learning of neural state-space models ​
Author: Bendeg'uz Gy"or"ok, Tam'as P'eni, Maarten Schoukens, Roland T'oth
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2607.17614v1 Announce Type: cross Abstract: Recent advances in deep-learning-based nonlinear system identification have led to encoder-based estimation of neural state-space (ANN-SS) models that achieve state-of-the-art performance in offline settings by estimating initial model states from pa...
270. Brain-Aligned Multi-Stream Video Transformers with Sparse Self-Selection ​
Author: Amir Hosein Fadaei, Mahyar Maleki, Mohammad-Reza A. Dehaqani
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.17625v1 Announce Type: cross Abstract: Modern video transformers typically ignore principles from primate vision and are rarely evaluated against neural data, limiting their biological interpretability. We introduce a sparse winner-takes-all token selection module that replaces dense self...
271. LFM: Leveraging Foundation Models for Source-Free Universal Domain Adaptation ​
Author: Jing Li, Pan Liu, Meng Zhao, Wanli Xue, Yanhong Yang, Xu Cheng, Fan Shi, Jianhua Zhang, Qinghua Hu, Shengyong Chen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.MM
arXiv:2607.17653v1 Announce Type: cross Abstract: Source-free universal domain adaptation (SF-UniDA) adapts a pre-trained source model to an unlabeled target domain under both covariate and label shifts, without access to source data. However, existing SF-UniDA methods rely on inefficient techniques...
272. An efficient adaptive dimension selection algorithm for multidimensional probit graded response models ​
Author: Yu Zhou, Yincai Tang, Bin Lv, Meng Gao
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2607.17654v1 Announce Type: cross Abstract: Multidimensional graded response models (MGRMs) are widely used for analyzing ordinal questionnaire data in psychological and educational assessments. A central challenge in applying these models is determining the number of latent dimensions. Conven...
273. Early Yield Prediction for Sugar Beet Fields using Satellite Data -- Learnings from Specialized Vision Transformers ​
Author: Philipp Vaeth, Bhumika Laxman Sadbhave, Denise Dejon, Gunther Schorcht, Magda Gregorova
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.17661v1 Announce Type: cross Abstract: Remote sensing has become an increasingly valuable tool for agricultural monitoring, particularly through the use of publicly available satellite imagery. However, effectively integrating domain knowledge into machine learning methods remains challen...
274. Equality, Equity, and Causality in Fairness Research: A Commentary on Cheng (2026) ​
Author: Youmi Suk
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML
arXiv:2607.17679v1 Announce Type: cross Abstract: This is an invited commentary on the Psychometrika focus article "Fairness Issues and Evaluation in Psychometrics and AI/ML: What Can We Learn from Each Field?" by Ying Cheng (2026, doi:10.1017/psy.2026.10110). Cheng offers a systematic comparison be...
275. An Adjoint-Sensitivity Framework for Lost-in-the-Middle Phenomena in Causal Residual Transformers ​
Author: Cheng Huan, Hongwei Yuan
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC
arXiv:2607.17696v1 Announce Type: cross Abstract: We develop an adjoint-sensitivity framework for positional influence in causal residual Transformers and separate unconditional analytic results from conditional boundary-shape conclusions. The principal unconditional theorem is the residual-to-depth...
276. Measuring Monosemanticity in Sparse Autoencoders via Latent Activation Coherence ​
Author: Katarzyna Filus, Sebastian Pokuci'nski
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2607.17770v1 Announce Type: cross Abstract: Within Explainable Artificial Intelligence, mechanistic interpretability uses Sparse Autoencoders (SAEs) to extract more interpretable features from neural representations. However, assessing their monosemanticity, and thus explanation quality, remai...
277. ETAS: An Effect-Typed Language for Agent Systems ​
Author: Huiri Tan, Yikun Wang, Puyang Zhang, Shangyu Li, Jiasi Shen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.PL, cs.AI, cs.LG, cs.MA
arXiv:2607.17780v1 Announce Type: cross Abstract: ETAS is a programming language for agent systems that treats model-backed agents, tool calls, prompts, typed memory, human approvals, policies, and execution traces as semantic program elements rather than library conventions. It separates determinis...
278. Reasoning as a Double-Edged Sword: Architecture and Cross-Stage Robustness in Vision-Language-Action Models ​
Author: Tuan Duong Trinh, Naveed Akhtar, Basim Azam
Published: 7/21/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CR, cs.LG
arXiv:2607.17786v1 Announce Type: cross Abstract: Does adding a reasoning step make a Vision-Language-Action (VLA) model more robust to perturbation? Intuitively, a policy that reasons before acting should absorb a perturbed input better than one that maps observations directly to actions. We test t...
279. Organization of computation in reservoir computing ​
Author: Mohab Abdalla, Damien Rontani
Published: 7/21/2026, 4:00:00 AM
Categories: cs.NE, cs.LG
arXiv:2607.17858v1 Announce Type: cross Abstract: Reservoir computing exploits nonlinear dynamical systems to encode temporal inputs into high-dimensional state space representations. Although reservoir performance is often characterized through memory, nonlinearity, and their tradeoff, such aggrega...
280. Entanglement geometry separates circuit cutting, classical hardness, and trainability ​
Author: Maria Gragera Garces, Sabina Dr\u{a}goi, Lirand"e Pira
Published: 7/21/2026, 4:00:00 AM
Categories: quant-ph, cs.DC, cs.ET, cs.LG
arXiv:2607.17872v1 Announce Type: cross Abstract: Circuit cutting promises to scale quantum computations beyond current hardware, but variational quantum advantage also requires low cutting overhead, classical hardness, and trainability. We show that these properties are strongly constrained by enta...
281. Beyond the Edge of Chaos: Stability-Expressivity Transfer in Reservoir Forecasting ​
Author: Yao Du, Xingang Wang
Published: 7/21/2026, 4:00:00 AM
Categories: nlin.CD, cs.LG, nlin.AO
arXiv:2607.17909v1 Announce Type: cross Abstract: The edge-of-chaos heuristic has long served as a guiding principle for designing reservoir computers, yet its relevance to machine performance remains elusive. Here, taking the spectral radius of the reservoir network as the control parameter, we sho...
282. Chemical filters for ultra-high-throughput materials screening and generation ​
Author: Kinga O. Mastej, Panyalak Detrattanawichai, Hyunsoo Park, Anthony Onwuli, Masahiro Negishi, Aron Walsh
Published: 7/21/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.AI, cs.LG
arXiv:2607.17910v1 Announce Type: cross Abstract: Generative artificial intelligence is rapidly transforming materials design by enabling de novo exploration of immense chemical spaces. Yet a large proportion of AI-generated compositions remain implausible, violating established chemical principles,...
283. AutoEncoder-Compressed Parallel Split Learning for Pre-trained Model Fine-Tuning ​
Author: Bas Meuwissen, Vasileios Tsouvalas, Nirvana Meratnia
Published: 7/21/2026, 4:00:00 AM
Categories: cs.DC, cs.LG
arXiv:2607.17913v1 Announce Type: cross Abstract: Distributed Fine-Tuning (DFT) of large-scale Foundation Models (FMs) on resource-constrained edge devices is limited by local compute constraints and communication overhead. Parallel Split Learning (PSL) reduces client-side computation by keeping few...
284. Value-Aware Prediction for Robust Multi-Agent Coordination Under Communication Loss ​
Author: Kemal Devrim Kafadar, Eren "Ozaltun, Mahmud Efnan \c{S}anl{\i}, Feyza Orak, Emirhan Gazi, Kubilay Ka\u{g}an K"om"urc"u, Naz{\i}m Kemal "Ure
Published: 7/21/2026, 4:00:00 AM
Categories: cs.MA, cs.LG, cs.RO
arXiv:2607.17914v1 Announce Type: cross Abstract: Robust multi-agent coordination relies heavily on inter-agent communication, which is frequently disrupted by physical and environmental constraints in real-world deployments. To maintain operation during these intermittent communication failures, ag...
285. PRIME: Plasticity Recovery in Multi-Agent Environments for UAV-Assisted Emergency Communication Networks ​
Author: Wen Qiu, Zhiqiang He, Wei Zhao, Hiroshi Masui
Published: 7/21/2026, 4:00:00 AM
Categories: cs.MA, cs.LG, cs.NI
arXiv:2607.17922v1 Announce Type: cross Abstract: Most reinforcement learning controllers for these networks assume stationary conditions, and the few that handle change react to the external environment while leaving the network's internal state unexamined. We show that sustained non-stationarity d...
286. Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization ​
Author: Zijian Zhao, Sen Li
Published: 7/21/2026, 4:00:00 AM
Categories: cs.MA, cs.LG
arXiv:2607.17924v1 Announce Type: cross Abstract: Multi-agent policy optimization, exemplified by PPO-based methods, is a key branch of cooperative Multi-Agent Reinforcement Learning (MARL). A central design question is how many neighboring agents\footnote{In this paper, "neighbors" refer not only t...
287. PAMD: Structured Adaptive Distances for Bisimulation Representations in Visual Reinforcement Learning ​
Author: Daegyeong Roh, Juho Bae, Han-Lim Choi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.18004v1 Announce Type: cross Abstract: Many visual reinforcement learning (RL) algorithms learn representations by matching latent distances to a behavioral distance induced by reward and transition similarity. In practice, the choice of the latent distance can strongly affect performance...
288. Remote Awareness of Seafloor Images Collected by AUVs over Low-Bandwidth Communication Links ​
Author: Adrian Bodenmann, Cailei Liang, Miquel Massot-Campos, Samuel Simmons, Alexander B. Phillips, Alberto Consensi, Matthew Kingsland, Rashiid Sherif, Stan Brown, Adam Riese, Blair Thornton
Published: 7/21/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, cs.RO, eess.IV
arXiv:2607.18013v1 Announce Type: cross Abstract: This paper introduces a method for real-time processing and transmission of autonomous underwater vehicle (AUV) imagery over low-bandwidth communication links. It leverages artificial intelligence (AI) techniques to identify a set of images that best...
289. Adaptive Adversaries: A Multi-Turn, Multi-LLM Benchmark for LLM Agent Security ​
Author: Devina Jain, David Hartmann, Chuan Li
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2607.18063v1 Announce Type: cross Abstract: LLM-based agents process external content, exposing them to prompt injection and multi-turn manipulation. Most safety benchmarks evaluate defenders against fixed attack pools collected before evaluation, single-turn or multi-turn. We present a 21-sce...
290. Hardware Mechanisms to Dynamically Throttle AI Performance ​
Author: Haiyue Ma, Lauren Malek, Joseph Forzani, David Wentzlaff
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2607.18069v1 Announce Type: cross Abstract: As more capable AI models are increasingly integrated into critical computer systems, the lack of control over AI intent motivates safety mechanisms. Existing software safeguards impose only behavioral constraints that can potentially be bypassed by ...
291. WorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Football Forecasting ​
Author: Zhaokai Wang, Tianlin Gui, Jiayuan Rao, Shangzhe Di, Yihong Tang, Dingli Liang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG
arXiv:2607.18084v1 Announce Type: cross Abstract: Predicting a football match before kickoff requires more than knowing past results: a model must use changing information and make a clear prediction before the answer is available. We present WorldCupArena, a dynamic benchmark for language models an...
292. SciForma: Structure-Faithful Generation of Scientific Diagrams ​
Author: Yuxuan Luo, Peng Zhang, Xinjie Zhang, Xun Guo, Zhouhui Lian, Yan Lu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.LG
arXiv:2607.18091v1 Announce Type: cross Abstract: Structural fidelity is essential to scientific methodology diagrams. To communicate research logic, these diagrams must faithfully render components, directional relations, and textual annotations. Since a single error, such as a reversed arrow or an...
293. How Does Alignment Tuning Shape Representations of Sycophancy and Related Cue-Induced Biases in LLMs? ​
Author: Prakhar Gupta, Terry Jingchen Zhang, Florent Draye, Bernhard Sch"olkopf, Zhijing Jin
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.18114v1 Announce Type: cross Abstract: Modern LLMs are alarmingly susceptible to surprisingly simple immaterial changes of input prompts: a casual hint, an incorrectly labeled few-shot example, or a fake prior assistant turn often flips an originally correct answer. We study where this su...
294. COVAriance-Induced Fairness Gap Penalty for Subgroup-Fair Clustering ​
Author: Kyungseon Lee, Hankyo Jeong, Kunwoong Kim, Kwanho Lee, Yongdai Kim
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.18119v1 Announce Type: cross Abstract: Fair clustering aims to make cluster assignments independent of sensitive attributes, but this goal becomes challenging when multiple sensitive attributes jointly define many subgroups. In such settings, directly extending existing fair clustering al...
295. ClouDens: Operational Context-Aware Anomaly Detection for Large-scale Cloud System Monitoring ​
Author: Thu T. H. Doan, Mohammad Saiful Islam, Andriy Miranskyy, Ngoc-Thanh Nguyen, Rogardt Heldal, Patrizio Pelliccione
Published: 7/21/2026, 4:00:00 AM
Categories: cs.NI, cs.LG
arXiv:2607.18127v1 Announce Type: cross Abstract: With the rapid growth of cloud computing infrastructures in scale and complexity, network monitoring for Large-scale Cloud Systems (LCSs) has become increasingly challenging, requiring automated and reliable anomaly detection to maintain service avai...
296. EVOLVE: Efficient Learned Volume Compression with Variable-Rate Encoding on a Cross-Domain Database ​
Author: Kaiyuan Tang, Maizhe Yang, Chaoli Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.GR, cs.DB, cs.LG
arXiv:2607.18187v1 Announce Type: cross Abstract: Large-scale scientific simulations generate volumetric data at rates that far outpace advances in storage and network bandwidth, making effective lossy compression increasingly critical. However, conventional compressors often struggle to preserve fi...
297. Certified Training for Convolutional Perturbations ​
Author: Benedikt Br"uckner, Alessio Lomuscio
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.18195v1 Announce Type: cross Abstract: Vision models have been found to be susceptible to perturbations such as motion blur induced at runtime by a shaking camera. This impedes their deployment in critical applications since phenomena such as slightly blurred vision might lead to failures...
298. PPL-Factory: Task-Aware and Budget-Aware Data Selection from Language Modeling to Reasoning ​
Author: Hang Zhang, Warren J. Gross
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.18199v1 Announce Type: cross Abstract: Not all training samples contribute equally to large language model fine-tuning. Selecting informative training samples can reduce the computational cost while preserving downstream performance. Many existing data selection methods rely on indirect h...
299. Unveiling Invariant and Transferable Latent Factors Across Heterogeneous Environments via ATLAS ​
Author: Yihong Gu, Katherine Liao, Tianxi Cai
Published: 7/21/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.ME, stat.ML, stat.TH
arXiv:2607.18209v1 Announce Type: cross Abstract: This paper considers a multi-environment factor model in which high-dimensional covariates are collected from heterogeneous environments, with auxiliary labels available in a subset of these environments. The joint distribution of the covariates may ...
300. Vector Search As Nearest Neighbor Matching: RAG-based Policy Learning in Causal Inference ​
Author: Masahiro Kato, Taka Kato
Published: 7/21/2026, 4:00:00 AM
Categories: econ.EM, cs.LG, math.ST, stat.ME, stat.ML, stat.TH
arXiv:2607.18225v1 Announce Type: cross Abstract: We propose one-step and two-step methods for policy learning with retrieval-augmented generation (RAG). We formulate RAG-based action selection under the potential outcome framework. In the two-step method, vector search retrieves action-specific nei...
301. Patch Policy: Efficient Embodied Control via Dense Visual Representations ​
Author: Gaoyue Zhou, Zichen Jeff Cui, Ada Langford, Bowen Tan, Yann LeCun, Lerrel Pinto
Published: 7/21/2026, 4:00:00 AM
Categories: cs.RO, cs.LG
arXiv:2607.18236v1 Announce Type: cross Abstract: Pretrained dense visual features from Vision Transformers (ViTs) are powerful yet have been underutilized in robot learning. Modern robot policies either compress each observation into a single global token, or rely on visual backbones trained from s...
302. The Many Senses of Visual Similarity: A Text-Prompted Image Perceptual Metric ​
Author: Sheng-Yu Wang, Yotam Nitzan, Aaron Hertzmann, Jun-Yan Zhu, Eli Shechtman, Alexei A. Efros, Richard Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.18237v1 Announce Type: cross Abstract: Human visual similarity judgments are context-dependent. For example, two images may be similar in shape but distinct in color. Existing perceptual similarity metrics, however, collapse these nuances into a single scalar value, offering no mechanism ...
303. Hierarchical Reinforcement Learning for Air Combat at DARPA's AlphaDogfight Trials ​
Author: Adrian P. Pope, Jaime S. Ide, Daria Micovic, Henry Diaz, David Rosenbluth, Lee Ritholtz, Jason C. Twedt, Thayne T. Walker, Kevin Alcedo, Daniel Javorsek
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2105.00990v3 Announce Type: replace Abstract: Autonomous control in high-dimensional, continuous state spaces is a persistent and important challenge in the fields of robotics and artificial intelligence. Because of high risk and complexity, the adoption of AI for autonomous combat systems has...
304. Ghosts in Neural Networks: Existence, Structure and Role of Infinite-Dimensional Null Space ​
Author: Sho Sonoda, Isao Ishikawa, Masahiro Ikeda
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2106.04770v2 Announce Type: replace Abstract: We study parameter nonuniqueness in continuous-width depth-two fully connected neural networks. Our main contribution is a direct method for solving the neural-network equation $S[\gamma]=f$. Starting from the Fourier expression of the synthesis op...
305. Automated Reinforcement Learning: An Overview ​
Author: Reza Refaei Afshar, Joaquin Vanschoren, Uzay Kaymak, Rui Zhang, Yaoxin Wu, Wen Song, Yingqian Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2201.05000v3 Announce Type: replace Abstract: Reinforcement Learning and, recently, Deep Reinforcement Learning are popular methods for solving sequential decision-making problems modeled as Markov Decision Processes. RL modeling of a problem and selecting algorithms and hyper-parameters requi...
306. Posts of Peril: Detecting Information About Hazards in Text ​
Author: Keith Burghardt, Daniel M. T. Fessler, Chyna Tang, Anne Pisor, Kristina Lerman
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2405.17838v3 Announce Type: replace Abstract: Socio-linguistic indicators of affectively-relevant phenomena, such as emotion or sentiment, are often extracted from text to better understand features of human-computer interactions, including on social media. However, an indicator that is often ...
307. A Survey of Features Used for Representing Black-box Single-objective Continuous Optimization ​
Author: Gjorgjina Cenikj, Ana Nikolikj, Ga\v{s}per Petelin, Niki van Stein, Carola Doerr, Tome Eftimov
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2406.06629v2 Announce Type: replace Abstract: This survey examines key advancements in designing features to represent optimization problem instances, algorithm instances, and their interactions within the context of single-objective continuous black-box optimization. These features support ma...
308. Regret Optimality of Sample Average Approximation for Data-Driven Newsvendor Problems: A General Optimization Perspective ​
Author: Jiameng Lyu, Shilin Yuan, Bingkun Zhou, Yuan Zhou
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, math.OC
arXiv:2407.04900v2 Announce Type: replace Abstract: Numerous existing studies have examined the performance of Sample Average Approximation (SAA) in the fundamental newsvendor problem. Despite these advances, critical gaps remain in two aspects. First, existing works focus on the linear-cost newsven...
309. Online-Score-Aided Federated Learning for Resource-Constrained Wireless Clients with Continual Data Arrival ​
Author: Ferdous Pervej, Minseok Choi, Andreas F. Molisch
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.NI, cs.SY, eess.SY
arXiv:2408.05886v5 Announce Type: replace Abstract: Heterogeneous system configurations of distributed clients connected to the central server (CS) via a time-varying wireless network pose significant challenges for popular distributed machine learning (ML) algorithms such as federated learning (FL)...
310. Distributed solar generation forecasting using attention-based deep neural networks for cloud movement prediction ​
Author: Maneesha Perera, Julian De Hoog, Kasun Bandara, Hansani Weeratunge, Saman Halgamuge
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2411.10921v2 Announce Type: replace Abstract: Accurate forecasts of distributed solar generation are necessary to maintain grid stability amid the increased uptake of distributed solar photovoltaic (PV) systems. However, the high variability of solar generation over short time intervals (secon...
311. BADTV: Unveiling Backdoor Threats in Third-Party Task Vectors ​
Author: Chia-Yi Hsu, Yu-Lin Tsai, Yu Zhe, Yan-Lun Chen, Chih-Hsun Lin, Chia-Mu Yu, Yang Zhang, Chun-Ying Huang, Jun Sakuma
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CR
arXiv:2501.02373v3 Announce Type: replace Abstract: Task arithmetic in large-scale pre-trained models enables agile adaptation to diverse downstream tasks without extensive retraining. By leveraging task vectors (TVs), users can perform modular updates through simple arithmetic operations like addit...
312. Mediator: Memory-efficient LLM Merging with Less Parameter Conflicts and Uncertainty Based Routing ​
Author: Kunfeng Lai, Zhenheng Tang, Xinglin Pan, Peijie Dong, Xiang Liu, Haolan Chen, Huacan Wang, Li Shen, Bo Li, Xiaowen Chu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2502.04411v3 Announce Type: replace Abstract: Model merging aggregates Large Language Models (LLMs) finetuned on different tasks into a stronger one. However, parameter conflicts between models leads to performance degradation in averaging. While model routing addresses this issue by selecting...
313. PLayer-FL: A Principled Approach to Personalized Layer-wise Cross-Silo Federated Learning ​
Author: Ahmed Elhussein, Florent Pollet, Gamze G"ursoy
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2502.08829v2 Announce Type: replace Abstract: Federated learning (FL) with non-IID data often degrades client performance below local training baselines. Partial FL addresses this by federating only early layers that learn transferable features, but existing methods rely on ad-hoc, architectur...
314. AdaGC: Enhancing LLM Pretraining Stability via Adaptive Gradient Clipping ​
Author: Guoxia Wang, Shuai Li, Congliang Chen, Jinle Zeng, Jiabin Yang, Dianhai Yu, Yanjun Ma, Li Shen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2502.11034v4 Announce Type: replace Abstract: Loss spikes remain a persistent obstacle in large-scale language model pretraining. While previous research has attempted to identify the root cause of loss spikes by investigating individual factors, we observe that, in practice, such spikes are t...
315. Supervised Reward Inference ​
Author: Will Schwarzer, Jordan Schneider, Philip S. Thomas, Scott Niekum
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2502.18447v3 Announce Type: replace Abstract: Existing approaches to reward inference typically assume that humans provide demonstrations according to specific behavior models. However, humans often indicate their goals through a wide range of behaviors, from actions that are suboptimal due to...
316. On the Interpolation Effect of Score Smoothing in Diffusion Models ​
Author: Zhengdao Chen
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML
arXiv:2502.19499v4 Announce Type: replace Abstract: Diffusion models have achieved remarkable progress in various domains with an intriguing ability to produce new data that do not exist in the training set. In this work, we study the hypothesis that such creativity arises from the neural network ba...
317. ExMAG: Learning of Maximally Ancestral Graphs ​
Author: Petr Ry\v{s}av'y, Pavel Ryt'i\v{r}, Xiaoyu He, Georgios Korpas, Jakub Mare\v{c}ek
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2503.08245v4 Announce Type: replace Abstract: In mixed graphs, there are both directed and bidirected edges. An extension of acyclicity to this mixed-graph setting is known as maximally ancestral graphs. This extension is of considerable interest in causal learning in the presence of confounde...
318. A Survey on Unlearnable Data ​
Author: Jiahao Li, Yiqiang Chen, Yunbing Xing, Yang Gu, Xiangyuan Lan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2503.23536v3 Announce Type: replace Abstract: Unlearnable data (ULD) has emerged as an innovative defense technique to prevent machine learning models from learning meaningful patterns from specific data, thus protecting data privacy and security. By introducing perturbations to the training d...
319. Potential failures of physics-informed machine learning in traffic flow modeling: theoretical and experimental analysis ​
Author: Yuan-Zheng Lei, Yaobang Gong, Dianwei Chen, Yao Cheng, Xianfeng Terry Yang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph
arXiv:2505.11491v4 Announce Type: replace Abstract: This study investigates why physics-informed machine learning (PIML) can fail in macroscopic traffic flow modeling. We define failure as cases where a PIML model underperforms both purely data-driven and purely physics-based baselines by a given th...
320. Prismatic Synthesis: Gradient-based Data Diversification Boosts Generalization in LLM Reasoning ​
Author: Jaehun Jung, Seungju Han, Ximing Lu, Skyler Hallinan, David Acuna, Shrimai Prabhumoye, Mostafa Patwary, Mohammad Shoeybi, Bryan Catanzaro, Yejin Choi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2505.20161v2 Announce Type: replace Abstract: Effective generalization in language models depends critically on the diversity of their training data. Yet existing diversity metrics often fall short of this goal, relying on surface-level heuristics that are decoupled from model behavior. This m...
321. Sign-SZPO: Provable Preference-based Reinforcement Learning with an Unknown Link Function ​
Author: Qining Zhang, Lei Ying
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2506.03066v2 Announce Type: replace Abstract: The link function, which characterizes the relationship between the preference for two trajectories and their returns, is a crucial component in designing RL algorithms that learn from preference feedback. Most existing methods, both theoretical an...
322. Can Interpretation Predict Behavior on Unseen Data? ​
Author: Victoria R. Li, Jenny Kaufmann, Tian Qin, Martin Wattenberg, David Alvarez-Melis, Naomi Saphra
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2507.06445v3 Announce Type: replace Abstract: Interpretability research often predicts model responses to targeted mechanistic interventions. But can we predict responses to unseen input data? We propose and demonstrate this alternate objective by using model internals to predict their out-of-...
323. Attentions Under the Microscope: A Comparative Study of Resource Utilization for Variants of Self-Attention ​
Author: Zhengyu Tian, Anantha Padmanaban Krishna Kumar, Hemant Krishnakumar, Reza Rawassizadeh
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE
arXiv:2507.07247v2 Announce Type: replace Abstract: As large language models (LLMs) and visual language models (VLMs) grow in scale and application, attention mechanisms have become a central computational bottleneck due to their high memory and time complexity. While many efficient attention varian...
324. Symmetric Behavior Regularized Policy Optimization ​
Author: Lingwei Zhu, Haseeb Shah, Zheng Chen, Martha White
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2508.04225v4 Announce Type: replace Abstract: Behavior Regularized Policy Optimization (BRPO) leverages asymmetric divergence regularization to mitigate distribution shift in offline reinforcement learning. This paper is the first to study the open question of symmetric BRPO. Using didactic ex...
325. Geometric origin of adversarial vulnerability in deep learning ​
Author: Yixiong Ren, Wenkang Du, Jianhui Zhou, Haiping Huang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.stat-mech, q-bio.NC
arXiv:2509.01235v2 Announce Type: replace Abstract: Balancing training accuracy and adversarial robustness has beeen a challenge since the birth of deep learning. Here, we introduce a geometry-aware deep learning framework that leverages layer-wise local training to sculpt the internal representatio...
326. NIRVANA: Structured Pruning Reimagined for Large Language Model Compression ​
Author: Mengting Ai, Tianxin Wei, Sirui Chen, Jingrui He
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2509.14230v2 Announce Type: replace Abstract: While structured pruning presents a highly effective pathway for accelerating Large Language Model (LLM) inference, existing methods frequently suffer from significant performance degradation and demand computationally retraining to recover capabil...
327. Multivariate Time Series Forecasting with Gate-Based Quantum Reservoir Computing on NISQ Hardware ​
Author: Wissal Hamhoum, Soumaya Cherkaoui, Jean-Frederic Laprade, Ola Ahmad, Shengrui Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.ET
arXiv:2510.13634v2 Announce Type: replace Abstract: Quantum reservoir computing (QRC) offers a hardware-friendly approach to temporal learning, yet most studies target univariate signals and overlook near-term hardware constraints. This work introduces a gate-based QRC for multivariate time series (...
328. Interpret Policies in Deep Reinforcement Learning using SILVER with RL-Guided Labeling: A Model-level Approach to High-dimensional and Multi-action Environments ​
Author: Yiyu Qian, Su Nguyen, Chao Chen, Qinyue Zhou, Liyuan Zhao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2510.19244v3 Announce Type: replace Abstract: Deep reinforcement learning (RL) achieves remarkable performance but lacks interpretability, limiting trust in policy behavior. The existing SILVER framework (Li, Siddique, and Cao 2025) explains RL policy via Shapley-based regression but remains r...
329. When Fewer Layers Break More Chains: Layer Pruning Harms Test-Time Scaling in LLMs ​
Author: Keyu Wang, Tian Lyu, Guinan Su, Lu Yin, Marco Canini, Jonas Geiping, Shiwei Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.22228v2 Announce Type: replace Abstract: Layer pruning has emerged as a widely adopted technique for improving the efficiency of large language models (LLMs). Although existing methods demonstrate strong performance retention on general knowledge tasks, their effect on long-chain reasonin...
330. BBOPlace-Bench: Benchmarking Black-Box Optimization for Chip Placement ​
Author: Ke Xue, Ruo-Tong Chen, Rong-Xi Tan, Xi Lin, Yunqi Shi, Siyuan Xu, Mingxuan Yuan, Chao Qian
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR, cs.NE
arXiv:2510.23472v2 Announce Type: replace Abstract: Chip placement is a vital stage in modern chip design, and black-box optimization (BBO) has been applied to it for decades. Early BBO efforts, however, were limited by immature problem formulations and inefficient algorithm designs, leading to wors...
331. InertialAR: Autoregressive 3D Molecule Generation with Inertial Frames ​
Author: Haorui Li, Weitao Du, Yuqiang Li, Hongyu Guo, Shengchao Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2510.27497v2 Announce Type: replace Abstract: Transformer-based autoregressive models have emerged as a unifying paradigm across modalities such as text and images, but their extension to 3D molecule generation remains underexplored. The gap stems from two fundamental challenges: (1) how to to...
332. ProtoTSNet: Interpretable Multivariate Time Series Classification With Prototypical Parts ​
Author: Bart{\l}omiej Ma{\l}kus, Szymon Bobek, Grzegorz J. Nalepa
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.02152v2 Announce Type: replace Abstract: Time series data is one of the most popular data modalities in critical domains such as industry and medicine. The demand for algorithms that not only exhibit high accuracy but also offer interpretability is crucial in such fields, as decisions mad...
333. Financial Management System for SMEs: Real-World Deployment of Accounts Receivable and Cash Flow Prediction ​
Author: Bart{\l}omiej Ma{\l}kus, Szymon Bobek, Grzegorz J. Nalepa
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2511.03631v2 Announce Type: replace Abstract: Small and Medium Enterprises (SMEs), particularly freelancers and early-stage businesses, face unique financial management challenges due to limited resources, small customer bases, and constrained data availability. This paper presents the develop...
334. ProDER: A Continual Learning Approach for Fault Prediction in Evolving Smart Grids ​
Author: Emad Efatinasab, Nahal Azadi, Davide Dalle Pezze, Gian Antonio Susto, Chuadhry Mujeeb Ahmed, Mirco Rampazzo
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2511.05420v2 Announce Type: replace Abstract: As smart grids evolve to meet growing energy demands and modern operational challenges, the ability to accurately predict faults becomes increasingly critical. However, existing AI-based fault prediction models struggle to ensure reliability in evo...
335. mHC-GNN: Manifold-Constrained Hyper-Connections for Graph Neural Networks ​
Author: Subhankar Mishra
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2601.02451v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) suffer from over-smoothing in deep architectures and expressiveness bounded by the 1-Weisfeiler-Leman (1-WL) test. We adapt Manifold-Constrained Hyper-Connections, recently proposed for Transformers, to graph neural net...
336. GlyRAG: Context-Aware Retrieval-Augmented Framework for Blood Glucose Forecasting ​
Author: Shovito Barua Soumma, Hassan Ghasemzadeh
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT
arXiv:2601.05353v2 Announce Type: replace Abstract: Accurate blood glucose forecasting using continuous glucose monitoring (CGM) data can support the early prediction of dysglycemic risk. However, current neural-network-based forecasting models treat CGM data as a purely numerical sequence without i...
337. Sparsity-Aware Low-Rank Representation for Efficient Fine-Tuning of Large Language Models ​
Author: Longteng Zhang, Sen Wu, Shuai Hou, Zhengyu Qing, Zhuo Zheng, Danning Ke, Qihong Lin, Qiang Wang, Shaohuai Shi, Xiaowen Chu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2601.16991v3 Announce Type: replace Abstract: Adapting large pre-trained language models to downstream tasks often entails fine-tuning millions of parameters or deploying costly dense weight updates, which hinders their use in resource-constrained environments. Low-rank Adaptation (LoRA) reduc...
338. Hybrid Mamba-Attention Neural Architecture for Channel Estimation ​
Author: Dianxin Luan, Chengsi Liang, Jie Huang, Zheng Lin, Kaitao Meng, John Thompson, Cheng-Xiang Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, eess.SP
arXiv:2601.17108v2 Announce Type: replace Abstract: This paper proposes a hybrid Mamba-attention neural architecture to achieve improved channel estimation for orthogonal frequency-division multiplexing (OFDM) waveforms, particularly for configurations with a large number of subcarriers. By integrat...
339. VAE with Hyperspherical Coordinates: Improving Anomaly Detection from Hypervolume-Compressed Latent Space ​
Author: Alejandro Ascarate, Leo Lebrat, Rodrigo Santa Cruz, Clinton Fookes, Olivier Salvado
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.18823v4 Announce Type: replace Abstract: Variational autoencoders (VAE) encode data into lower-dimensional latent vectors before decoding those vectors back to data. Once trained, one can hope to detect out-of-distribution (abnormal) latent vectors, but several issues arise when the laten...
340. Tabular Foundation Models Can Do Survival Analysis ​
Author: Da In Kim, Wei Siang Lai, Kelly W. Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2601.22259v2 Announce Type: replace Abstract: While tabular foundation models have achieved remarkable success in classification and regression, adapting them to model time-to-event outcomes for survival analysis is non-trivial due to right-censoring, where data observations may end before the...
341. Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals ​
Author: Zihan Dong, Zhixian Zhang, Yang Zhou, Can Jin, Ruijia Wu, Linjun Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.ST, stat.ME, stat.ML, stat.TH
arXiv:2602.03061v2 Announce Type: replace Abstract: Evaluating mathematical reasoning in LLMs is constrained by limited benchmark sizes and inherent model stochasticity, yielding high-variance accuracy estimates and unstable rankings across platforms. On difficult problems, an LLM may fail to produc...
342. UnMaskFork: Test-Time Scaling for Masked Diffusion via Deterministic Action Branching ​
Author: Kou Misaki, Takuya Akiba
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2602.04344v2 Announce Type: replace Abstract: Test-time scaling strategies have effectively leveraged inference-time compute to enhance the reasoning abilities of Autoregressive Large Language Models. In this work, we demonstrate that Masked Diffusion Language Models (MDLMs) are inherently ame...
343. Thermodynamic Limits of Physical Intelligence ​
Author: Koichi Takahashi, Yusuke Hayashi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT
arXiv:2602.05463v2 Announce Type: replace Abstract: Modern AI systems achieve remarkable capabilities at the cost of substantial energy consumption. To connect intelligence to physical efficiency, we propose two complementary bits-per-joule metrics under explicit accounting conventions: (1) Thermody...
344. When Benchmarks Lie: Evaluating Malicious Prompt Classifiers Under True Distribution Shift ​
Author: Max Fomin
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.14161v2 Announce Type: replace Abstract: Detecting prompt injection, jailbreak attacks, and harmful requests is critical for deploying LLM-based agents safely, yet current evaluation practices in this literature overestimate generalization. We train activation-based classifiers (linear pr...
345. Bigger Is Safer: Provable Robustness in In-Context Learning Scales with Capacity ​
Author: Di Zhang, Ningxu Zhang, Zimeng Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, stat.ML
arXiv:2602.17743v2 Announce Type: replace Abstract: In-context learning (ICL) allows large language models to adapt to new tasks from a few examples without updating their parameters. Existing theories explain ICL by assuming the test task distribution matches pretraining -- an assumption that break...
346. Audio-Visual Continual Test-Time Adaptation without Forgetting ​
Author: Sarthak Kumar Maharana, Akshay Mehra, Bhavya Ramakrishna, Yunhui Guo, Guan-Ming Su
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.SD
arXiv:2602.18528v4 Announce Type: replace Abstract: Audio-visual continual test-time adaptation involves continually adapting a source audio-visual model at test-time, to unlabeled non-stationary domains, where either or both modalities can be distributionally shifted, which hampers online cross-mod...
347. Rethinking Efficiency in Neural Combinatorial Optimization: Batched Preference Optimization with Mamba ​
Author: Zhenxing Xu, Zeyuan Ma, Weidong Bao, Yan Zheng, Chongshuang Hu, Ji Wang, Zhiguang Cao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.20730v3 Announce Type: replace Abstract: We study efficiency as a first-class objective in Neural Combinatorial Optimization (NCO) and present ECO, an efficient learning framework that combines batched preference optimization with a Mamba backbone. Instead of tightly interleaving every po...
348. GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning ​
Author: Ningyuan Yang, Weihua Du, Weiwei Sun, Sean Welleck, Yiming Yang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL
arXiv:2602.21492v2 Announce Type: replace Abstract: Reinforcement learning (RL) has become a central post-training paradigm for large language models (LLMs), but its performance is highly sensitive to the quality of training problems. This sensitivity stems from the non-stationarity of RL: rollouts ...
349. Routing by Reaching: Composition of Pre-trained GFlowNets for Multi-Objective Generation ​
Author: Seokwon Yoon, Youngbin Choi, Seunghyuk Cho, Seungbeom Lee, MoonJeong Park, Dongwoo Kim
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2602.21565v3 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) learn to sample diverse candidates in proportion to a reward function, making them well-suited for scientific discovery, where exploring multiple promising solutions is crucial. Further extending GFlowNets to mu...
350. Induction Meets Biology: Mechanisms of Repeat Detection in Protein Language Models ​
Author: Gal Pomerants, Yaniv Nikankin, Anja Reusch, Tomer Tsaban, Ora Schueler-Furman, Yonatan Belinkov
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM
arXiv:2602.23179v5 Announce Type: replace Abstract: Protein sequences are abundant in repeating segments, both as exact copies and as approximate segments with mutations. These repeats are important for protein structure and function, motivating decades of algorithmic work on repeat identification. ...
351. Long Range Frequency Tuning for QML ​
Author: Michael Poppel, Markus Baumann, Sebastian W"olckert, Claudia Linnhoff-Popien, Jonas Stein
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET, quant-ph
arXiv:2602.23409v3 Announce Type: replace Abstract: Angle-encoded variational quantum circuits admit a truncated Fourier series representation of their output, but approximating functions with maximum frequency $\omega_{\max}$ using fixed unary encoding requires $\mathcal{O}(\omega_{\max})$ encoding...
352. Breaking the Factorization Barrier in Diffusion Language Models ​
Author: Ian Li, Zilei Shao, Benjie Wang, Rose Yu, Guy Van den Broeck, Anji Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.00045v3 Announce Type: replace Abstract: Diffusion language models theoretically allow for efficient parallel generation but are practically hindered by the ``factorization barrier'': the assumption that simultaneously predicted tokens are independent. This limitation forces a trade-off: ...
353. Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training ​
Author: Xi Wang, Wenbo Lu, Shengjie Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.00454v3 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) enable fine-tuning large language models to approximate reward-proportional posteriors, but they remain prone to mode collapse, manifesting as prefix collapse and length bias. We attribute this to two factors: (...
354. From Heuristic Selection to Automated Algorithm Design: LLMs Benefit from Strong Priors ​
Author: Qi Huang, Furong Ye, Ananta Shahane, Thomas B"ack, Niki van Stein
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.NE
arXiv:2603.02792v2 Announce Type: replace Abstract: Large Language Models (LLMs) have already been widely adopted for automated algorithm design, demonstrating strong abilities in generating and evolving algorithms across various fields. Existing work has largely focused on examining their effective...
355. IoUCert: Robustness Verification for Anchor-based Object Detectors ​
Author: Benedikt Br"uckner, Alejandro J. Mercado, Yanghao Zhang, Panagiotis Kouvaros, Alessio Lomuscio
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR, cs.CV
arXiv:2603.03043v3 Announce Type: replace Abstract: While formal robustness verification has seen significant success in image classification, scaling these guarantees to object detection remains notoriously difficult due to complex non-linear coordinate transformations and Intersection-over-Union (...
356. Preventing Learning Stagnation in PPO by Scaling to 1 Million Parallel Environments ​
Author: Michael Beukman, Khimya Khetarpal, Zeyu Zheng, Will Dabney, Jakob Foerster, Michael Dennis, Clare Lyle
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.06009v2 Announce Type: replace Abstract: An agent's performance stagnating at a suboptimal level is a common problem in deep on-policy RL. Focusing on PPO, we show that plateaus in certain regimes arise not because of known exploration, capacity, or optimisation challenges, but because sa...
357. DirPA: Addressing Prior Shift in Imbalanced Few-shot Crop-type Classification ​
Author: Joana Reuss, Ekaterina Gikalo, Marco K"orner
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CV
arXiv:2603.12905v2 Announce Type: replace Abstract: Real-world agricultural monitoring is often hampered by severe class imbalance and high label acquisition costs, resulting in significant data scarcity. In few-shot learning (FSL) -- a framework specifically designed for data-scarce settings -- , t...
358. L2GTX: From Local to Global Time Series Explanations ​
Author: Ephrem Tibebe Mekonnen, Luca Longo, Lucas Rizzo, Pierpaolo Dondio
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2603.13065v2 Announce Type: replace Abstract: Deep learning models achieve high accuracy in time series classification, yet understanding their class-level decision behaviour remains challenging. Explanations for time series must respect temporal dependencies and identify patterns that recur a...
359. Time-Aware Prior Fitted Networks for Zero-Shot Forecasting with Exogenous Variables ​
Author: Andres Potapczynski, Ravi Kiran Selvam, Tatiana Konstantinova, Malcolm Wolff, Kin G. Olivares, Ruijun Ma, Michael W. Mahoney, Andrew Gordon Wilson, Boris N. Oreshkin, Dmitry Efimov
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.15802v2 Announce Type: replace Abstract: In many time series forecasting settings, the target time series is accompanied by exogenous covariates, such as promotions and prices in retail demand; temperature in energy load; calendar and holiday indicators for traffic or sales; and grid load...
360. Stochastic Resetting Accelerates Reinforcement Learning Beyond Random Search ​
Author: Jello Zhou, David J. Schwab, Vudtiwat Ngampruetikorn
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn, cond-mat.stat-mech, cs.SY, eess.SY, physics.bio-ph
arXiv:2603.16842v2 Announce Type: replace Abstract: Stochastic resetting -- intermittently returning a process to a fixed reference state -- has emerged as an effective mechanism for optimizing first-passage properties. Existing theory largely treats processes that search but do not learn: the searc...
361. NanoZK: Privacy-Preserving Verifiable Inference for Large Language Models via Layerwise Zero-Knowledge Proofs ​
Author: Zhaohui Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2603.18046v2 Announce Type: replace Abstract: We present NanoZK, a zero-knowledge proof system for verifiable LLM inference: clients and third-party auditors check that a provider executed the advertised model on a committed input without learning weights or activations. NanoZK introduces a la...
362. SkillRouter: Skill Routing for LLM Agents at Scale ​
Author: YanZhao Zheng, ZhenTao Zhang, Chao Ma, YuanQiang Yu, JiHuai Zhu, Yong Wu, Tianze Xu, Baohua Dong, Hangcheng Zhu, Ruohui Huang, Gang Yu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.22455v5 Announce Type: replace Abstract: Reusable skills let LLM agents package task-specific procedures, tool affordances, and execution guidance into modular building blocks. As skill ecosystems grow to tens of thousands of entries, exposing every skill at inference time becomes infeasi...
363. Precision-Varying Prediction (PVP): Robustifying ASR systems against adversarial attacks ​
Author: Mat'ias Pizarro, Raghavan Narasimhan, Jonas Killian, Asja Fischer
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, eess.AS
arXiv:2603.22590v2 Announce Type: replace Abstract: With the increasing deployment of automated and agentic systems, ensuring the adversarial robustness of automatic speech recognition (ASR) models has become highly relevant. We observe that changing the precision of an ASR model during inference re...
364. SpecXMaster Technical Report ​
Author: Yutang Ge, Yaning Cui, Hanzheng Li, Jun-Jie Wang, Fanjie Xu, Jinhan Dong, Yongqi Jin, Dongxu Cui, Peng Jin, Guojiang Zhao, Hengxing Cai, Tianci Yangfeng, Xueqing Chen, Hongshuai Wang, Rong Zhu, Linfeng Zhang, Xiaohong Ji, Zhifeng Gao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.23101v3 Announce Type: replace Abstract: Intelligent spectroscopy serves as a pivotal element in AI-driven closed-loop scientific discovery, functioning as the critical bridge between matter structure and artificial intelligence. However, conventional expert-dependent spectral interpretat...
365. Stochastic Dimension Zeroth-Order Estimator: Stable and Memory-Efficient Training of PINNs ​
Author: Zhangyong Liang, Huanhuan Gao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2603.24002v3 Announce Type: replace Abstract: Physics-Informed Neural Networks (PINNs) for high-dimensional and high-order partial differential equations (PDEs) are primarily constrained by the $\mathcal{O}(d^k)$ spatial derivative complexity and the $\mathcal{O}(P)$ memory overhead of backpro...
366. Neural Global Optimization via Iterative Refinement from Noisy Samples ​
Author: Qusay Muzaffar, David Levin, Michael Werman
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2604.03614v2 Announce Type: replace Abstract: Global optimization of black-box functions from noisy samples is a fundamental challenge in machine learning and scientific computing. Traditional methods such as Bayesian Optimization often converge to local minima on multi-modal functions, while ...
367. DIB-OD: Preserving the Invariant Core for Robust Heterogeneous Graph Adaptation via Decoupled Information Bottleneck and Online Distillation ​
Author: Yang Yan, Yunxuan Li, Qiuyan Wang, Tianjin Huang, Qiudong Yu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2604.10882v2 Announce Type: replace Abstract: Graph Neural Network pretraining is pivotal for leveraging unlabeled graph data. However, generalizing across heterogeneous domains remains a major challenge due to severe distribution shifts. Existing methods primarily focus on intra-domain patter...
368. Frequency-Corrupt Based Graph Self-Supervised Learning ​
Author: Haojie Li, Mengjiao Zhang, Guanfeng Liu, Qiang Hu, Yan Wang, Junwei Du
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.SI
arXiv:2604.15699v2 Announce Type: replace Abstract: Graph self-supervised learning can reduce the need for labeled graph data and has been widely used in recommendation, social networks, and other web applications. However, existing methods often underuse high-frequency signals and may overfit to sp...
369. Sketching the Readout of Large Language Models for Scalable Data Attribution and Valuation ​
Author: Yide Ran, Jianwen Xie, Minghui Wang, Wenjin Zheng, Denghui Zhang, Chuan Li, Zhaozhuo Xu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.16197v2 Announce Type: replace Abstract: Data attribution and valuation are critical for understanding data-model synergy for Large Language Models (LLMs), yet existing gradient-based methods suffer from scalability challenges on LLMs. Inspired by human cognition, where decision making re...
370. Preserving Clusters in Error-Bounded Lossy Compression of Scientific Particle Data ​
Author: Congrong Ren, Sheng Di, Katrin Heitmann, Franck Cappello, Hanqi Guo
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2604.18801v2 Announce Type: replace Abstract: Scientific particle simulations in cosmology, molecular dynamics, and fluid dynamics produce large-scale datasets whose storage, movement, and analysis increasingly rely on lossy compression. However, existing compressors typically bound only point...
371. Generative Augmentation of Imbalanced Flight Records for Flight Diversion Prediction: A Multi-objective Optimisation Framework ​
Author: Karim Aly, Alexei Sharpanskykh, Jacco Hoekstra
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.20288v2 Announce Type: replace Abstract: Flight diversions are rare but high-impact events in aviation, making their reliable prediction vital for both safety and operational efficiency. However, their scarcity in historical records impedes the training of machine learning models used to ...
372. FETS Benchmark: Foundation Models Enable Scalable and Generalizable Energy Time Series Forecasting ​
Author: Marco Obermeier, Marco Pruckner, Florian Haselbeck, Andreas Zeiselmair
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE
arXiv:2604.22328v2 Announce Type: replace Abstract: Driven by the transition towards a climate-neutral energy system, accurate energy time series forecasting is critical for planning and operations. Yet, it remains a dataset-specific task, requiring comprehensive training data, limiting scalability,...
373. DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training ​
Author: Tianhao Hu, Xiangcheng Liu, Yuchun Miao, Youshao Xiao, Hongyu Zang, Yang Zheng, Xuan Huang, Jinrui Ding, Yufei Zhang, Yu Yang, Yi-Kai Zhang, Yueqing Sun, Chengcheng Han, Xiandi Ma, Wei Wang, Qi Gu, Yerui Sun, Yuchen Xie, Xunliang Cai
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2604.26256v2 Announce Type: replace Abstract: Reinforcement learning (RL) has become a critical paradigm for LLM post-training, yet the rollout phase -- accounting for 50--80% of total step time -- is bottlenecked by skewed generation: long-tailed trajectories indispensable for model performan...
374. Differentiable latent structure discovery for interpretable forecasting in clinical time series ​
Author: Ivan Lerner, Jean Feydy, Alexandre Kalimouttou, Anita Burgun, Francis Bach
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2604.27967v2 Announce Type: replace Abstract: Background: We introduce StructGP, a continuous-time multi-task Gaussian process that couples process convolutions with differentiable structure learning to uncover a sparse, ordered directed acyclic graph (DAG) of inter-variable dependencies while...
375. How Many Iterations to Jailbreak? Dynamic Budget Allocation for Multi-Turn LLM Evaluation ​
Author: Shai Feldman, Yaniv Romano
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.06605v5 Announce Type: replace Abstract: Evaluating and predicting the performance of large language models (LLMs) in multi-turn conversational settings is critical yet computationally expensive; key events -- e.g., jailbreaks or successful task completion by an agent -- often emerge only...
376. A Systematic Investigation of RL-Jailbreaking in LLMs ​
Author: Montaser Mohammedalamen, Kevin Roice, Reginald McLean, Alyssa Lefaivre \v{S}kopac
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.07032v3 Announce Type: replace Abstract: The evolution of generative models from next-token predictors to autonomous engines of complex systems necessitates rigorous safety hardening. Adversarial jailbreaking, the strategic manipulation of models to elicit harmful output, remains a primar...
377. AQKA: Active Quantum Kernel Acquisition Under a Shot Budget ​
Author: Jian Xu, Chao Li, Delu Zeng, John Paisley, Qibin Zhao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.14672v2 Announce Type: replace Abstract: Estimating an $N \times N$ quantum kernel from circuit fidelities requires $\Theta(N^2 S)$ measurement shots, the dominant bottleneck for deployment on near-term hardware. Existing budget-saving methods (Nystr"om-QKE, ShoFaR, kernel-target alignme...
378. Every Component is a Lookup: Token Attribution and Composition from a Single Decomposition ​
Author: Po-Kai Chen, Aske Plaat, Niki van Stein
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2605.23393v2 Announce Type: replace Abstract: Mechanistic interpretability of transformers requires identifying not just which components matter but how they compose into the computational route that produced a prediction. Both attention and MLP follow a shared key-value template $\phi(S)U$. W...
379. The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible ​
Author: Lauri Lov'en, Nam Do, Hassan Mehmood, Dinesh Kumar Sah, Sasu Tarkoma
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.GT, stat.ML
arXiv:2605.25739v2 Announce Type: replace Abstract: We prove that no reinforcement learning policy with confidence-gated autonomy can simultaneously achieve maximum helpfulness, optimal calibration, and full autonomy under rational oversight, whenever some tasks exceed the agent's reliable competenc...
380. AOE: Exhaustive Out-of-Distribution Detection via Recalibrating Outlier Labels ​
Author: Fengqiang Wan, Qing-Yuan Jiang, Fu Shen, Yang Yang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2605.28021v2 Announce Type: replace Abstract: Out-of-distribution (OOD) detection is essential for deploying machine learning models in open-world and safety-critical scenarios, where test inputs may deviate from the training distribution and overconfident predictions on unknown samples can le...
381. How Much Is a Dataset Worth? Scaling Laws, the Vendi Score, and Matrix Spectral Functions ​
Author: Jeff A. Bilmes, Gantavya Bhatt, Arnav M. Das
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.IT, math.IT
arXiv:2605.29448v2 Announce Type: replace Abstract: Neural scaling laws appraise data through dataset size, while the Vendi Score uses quantum entropy to measure dataset value. We show both that common neural-scaling-law objectives and the Vendi Score are submodular. We further show that the Vendi S...
382. Singularity-aware Optimization via Randomized Geometric Probing: Towards Stable Non-smooth Optimization ​
Author: Ruoran Xu, Borong She, Xiaobo Jin, Qiufeng Wang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC
arXiv:2605.29547v2 Announce Type: replace Abstract: Deep learning optimization relies heavily on the assumption of smooth loss landscapes, a condition systematically violated by modern architectures due to non-smooth components such as ReLU activations and quantization operators. In such non-smooth ...
383. Autoregressive Diffusion World Models for Off-Policy Evaluation of LLM Agents ​
Author: Kaixuan Liu, Guojun Xiong, Weinan Zhang, Shengpu Tang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.05558v2 Announce Type: replace Abstract: Evaluating large language model (LLM) agents in multi-turn interactive environments is expensive and risky, as it requires online environment interaction. We propose ADWM (Autoregressive Diffusion World Model), an evaluation framework that estimate...
384. FAIR-Calib: Frontier-Aware Instability-Reweighted Calibration for Post-Training Quantization of Diffusion Large Language Models ​
Author: Haoyu Huang, Linlin Yang, Sheng Xu, Boyu Liu, Guodong Guo, Zhongqian Fu, Hang Zhou, Baochang Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.06547v2 Announce Type: replace Abstract: Diffusion Large Language Models (dLLMs) refine tokens iteratively but commit them irreversibly, leading to a "stability lag" where early decisions remain fragile even after being written. We reveal that Post-Training Quantization (PTQ) error easily...
385. SPIN: Decentralized Swarm Control via Tensorized Policy Coordination ​
Author: Zhaowen Fan, Yunxiang Han
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.MA, cs.SI
arXiv:2606.07557v2 Announce Type: replace Abstract: Decentralized multi-agent swarm coordination remains fundamentally challenged by the combinatorial scaling of joint action spaces and high-overhead training or optimization loops when managing localized group behaviors. To address this problem from...
386. FlashMemory-DeepSeek-V4: Lightning Index Ultra-Long Context via Lookahead Sparse Attention ​
Author: Yan Wang, Qifan Zhang, Jiachen Yu, Tian Liang, Dongyang Ma, Xiang Hu, Zibo Lin, Chunyang Li, Zhichao Wang, Miao Peng, Nuo Chen, Jia Li, Yujiu Yang, Haitao Mi, Dong Yu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.09079v3 Announce Type: replace Abstract: Conventional LLMs keep the full KV cache loaded during decoding, causing a severe GPU memory bottleneck for ultra-long context serving. In this report, we propose \textbf{Lookahead Sparse Attention (LSA)}, a novel inference paradigm powered by a Ne...
387. RoVE: Rotary Value Embeddings Attention for Relative Position-dependent Value Pathways ​
Author: Alejandro Garc'ia-Castellanos, Maurice Weiler, Erik J Bekkers
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.11275v2 Announce Type: replace Abstract: Rotary Position Embeddings (RoPE) make attention scores position-relative but leave the value pathway position-blind: the message sent by a value token is the same regardless of its distance from the query. We propose RoVE, a parameter-free modific...
388. Interpretable Factor Decomposition for Decision Intelligence in Large-Scale Financial Markets: Evidence from China's A-Share Market ​
Author: Xiao Han, Yao Xiao, Zhen Zhang, Moxuan Zheng
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CE
arXiv:2606.12843v2 Announce Type: replace Abstract: We present an interpretable machine learning pipeline to decompose cross-sectional equity return predictability into auditable factor contributions. We apply an XGBoost model with TreeSHAP attribution and conduct stress testing on 3,632 Chinese A-s...
389. Overlapping Schwarz Attention: Hierarchical Attention via Domain Decomposition ​
Author: Stephan K"ohler, Oliver Rheinbach
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2606.18525v2 Announce Type: replace Abstract: We propose a hierarchical attention mechanism based on two-level overlapping Schwarz domain decomposition. The method is motivated by domain decomposition methods in partial differential equations which combine local subdomain corrections with a co...
390. Short-Term Electricity Demand Forecasting for New England: A Comprehensive Machine Learning Benchmark with Weather, Calendar, and COVID-19 Indicators ​
Author: Reza Ghanavati, Behrooz Mosallaei
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY
arXiv:2606.20918v2 Announce Type: replace Abstract: Accurate short-term electricity demand forecasting is critical for reliable power system operation, energy market planning, and infrastructure optimization. This paper benchmarks ten machine learning models for daily electricity demand forecasting ...
391. Right Knowledge, Wrong Answer: Characterizing Parametric Temporal Conflict in Open-Weight Language Models ​
Author: Elias Hossain, Sourav Saha, Tasfia Nuzhat Ornee, Sanjeda Sara Jennifer, Umesh Chandra Biswas, Shubhashis Roy Dipta, Rajib Rana, Niloofar Yousefi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.CL
arXiv:2606.20959v2 Announce Type: replace Abstract: Language models may encode both outdated facts and their newer replacements. We introduce Parametric Temporal Conflict (PTC), where the newer fact is present and recoverable, but the default forward pass prefers the outdated one. We release a deter...
392. TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning ​
Author: Yuanda Xu, Zhengze Zhou, Hejian Sang, Xiaomin Li, Jiaxin Zhang, Xinchen Du, Sen Na, Zhipeng Wang, Alborz Geramifard
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2606.32017v3 Announce Type: replace Abstract: Agentic reinforcement learning requires assigning credit to environment-facing actions such as searches, clicks, edits, navigation commands, and object interactions. Standard GRPO uses the final verifier outcome as a uniform advantage over all acti...
393. When Can You Debias an LLM Judge? Identifiability Limits, a Test, and Designs for Top-k Ranking ​
Author: Jian Xu, Delu Zeng, John Paisley, Qibin Zhao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.02104v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as cheap, scalable judges that compare candidate outputs pairwise. Because such judges prefer verbose or well-formatted answers, the natural fix is to add bias covariates to a Bradley--Terry model ...
394. Tensor-Train Joint Modeling for Few-Step Discrete Diffusion ​
Author: Byoungkwon Kim, Minhyuk Sung
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.03788v2 Announce Type: replace Abstract: Discrete diffusion promises orders-of-magnitude faster generation than autoregressive (AR) models for sequential discrete data, yet its full potential of few-step generation has remained out of reach due to a fundamental structural limitation. The ...
395. No Subspace to Track: Non-Identifiability and Optimizer State in Low-Rank Training ​
Author: Noel Thomas
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML
arXiv:2607.05872v2 Announce Type: replace Abstract: Memory-efficient optimizers such as GaLore train large language models by projecting gradients onto a rank-r subspace recomputed every T steps, assuming this subspace is a slowly drifting object that can be tracked. We show that beyond a small repr...
396. LLM-Guided Task-Semantic Field Factorization for Industrial Process Forecasting ​
Author: Youcheng Zong, Runda Jia, Mingxuan Ren, Dakuo He
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.SY, eess.SY
arXiv:2607.06623v2 Announce Type: replace Abstract: Process industries rely on time-series forecasting and soft sensing to estimate quality variables that are hard to measure online. Labeled data are scarce, operating regimes change frequently, and retraining models or rebuilding alignment pipelines...
397. Nonlinear Bandit ​
Author: Tianshuo Zheng, Ting Wu, Zhi-Hua Zhou, Keqin Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.07304v2 Announce Type: replace Abstract: In this paper we first study the problem of generalized linear bandit (GLB) under heavy-tailed noise. The characteristics of heavy-tailed distributions are widely observed in real-world applications such as personalized recommendation, financial ma...
398. TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning ​
Author: Fangxu Yu, Tao Feng, Dehai Min, Lu Cheng, Ge Liu, Tianyi Zhou
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.08940v2 Announce Type: replace Abstract: Time series reasoning is essential for real-world problem-solving. While both Large Language Models (LLMs) and Vision-Language Models (VLMs) can reason about time-series data, their capabilities are complementary: LLMs process time series as text s...
399. Metadata-Free Meta-Reweighted Direct Preference Optimization under Noisy Preference Labels ​
Author: Hua Qu, Yifan Li, Xiaodong Yuan
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.09796v2 Announce Type: replace Abstract: Direct Preference Optimization (DPO) has become an important method for aligning large language models (LLMs) with human preferences because it removes the need for explicit reward modeling and reinforcement learning. However, its performance depen...
400. Auditing Invisible Weight Updates with Reference Traces ​
Author: Zekai Shang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG
arXiv:2607.09800v2 Announce Type: replace Abstract: Master weights and stochastic rounding bypass invisible stored-weight updates but do not locate lost direct-storage proposals or parameters worth protecting. We ask what one matched high-precision pilot trace can audit before a low-precision run. T...
401. LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning ​
Author: Ning Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.10139v2 Announce Type: replace Abstract: Selecting the correct answer from a pool of candidate reasoning chains is the engine of test-time scaling, yet the standard selectors each carry a cost: self-consistency inherits the errors of the single model it resamples, and trained reward model...
402. Adaptive Compute in Latent World Models: When Depth Helps, Hurts, or Doesn't Matter ​
Author: Achyuthan Sivasankar
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.10203v3 Announce Type: replace Abstract: Adaptive compute for world models -- early-exit or mixture-of-depths predictors that spend variable depth per rollout step -- presumes that extra depth buys better predictions. In autoregressive rollouts, where planning actually happens, that premi...
403. Exact and Certified Data Shapley for Weighted k-Nearest-Neighbor Regression and Soft-Label Prediction ​
Author: Zongye Lyu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DS
arXiv:2607.11956v2 Announce Type: replace Abstract: Data Shapley answers which training points are worth what, and its nearest-neighbor specialization is the version actually deployed, shipped by toolkits such as pyDVL and OpenDataVal. Exact algorithms exist for unweighted nearest-neighbor classific...
404. Tabular Foundation Models for Discrete Choice Estimation ​
Author: Liu Liu, Dan Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, econ.EM
arXiv:2607.13314v2 Announce Type: replace Abstract: Tabular foundation models (TFMs) generate predictions on structured data via in-context learning, without task-specific estimation. We ask whether TFMs can be effectively applied to discrete choice, a central demand estimation framework in marketin...
405. Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values ​
Author: Jan Betley, Johannes Treutlein, Jan Dubi'nski, Harry Mayne, Karol Ga{\l}\k{a}zka, Niels Warncke, Anna Sztyber-Betley, Owain Evans
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR
arXiv:2607.14345v3 Announce Type: replace Abstract: People use language models for practical questions whose answers are difficult to verify. We show that models exhibit covert value leakage: the information they provide is influenced by their own values, without this influence being disclosed to th...
406. LATTICE: Graph Self-Supervised Learning for Multimodal Spatial Omics Integration ​
Author: Jagan Mohan Reddy Dwarampudi, Veena Kochat, Suresh Satpati, Hien Van Nguyen, Kunal Rai, Tania Banerjee
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, q-bio.GN
arXiv:2607.14410v2 Announce Type: replace Abstract: Spatially resolved omics studies increasingly combine transcriptomic and epigenomic assays, yet downstream analysis is often still performed using single-modality pipelines. We present LATTICE (Latent Alignment of Tissue-level and Transcriptomic In...
407. LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget ​
Author: Changhai Zhou, Kieran Liu, Yuhua Zhou, Qian Qiao, Jun Gao, Harry Zhang, Irvine Lu, Nolan Ho, Lucian Li, Andrew Lei, Cleon Cheng, Steven Chiang, Yihang Zeng, Di Zhang, Rio Yang, Kaijie Chen, Andrew Chen, Pony Ma, Weizhong Zhang, Cheng Jin
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.DC
arXiv:2607.14952v2 Announce Type: replace Abstract: A widening gap separates million-token inference from RL post-training, which remains at 256K tokens or below. The gap matters for AI agents, whose observations, tool outputs, documents, and decisions accumulate over long trajectories. Unlike infer...
408. When Does Muon Help Agentic Reinforcement Learning? ​
Author: Kai Ruan, Jinghao Lin, Zihe Huang, Ziqi Zhou, Qianshan Wei, Xuan Wang, Hao Sun
Published: 7/21/2026, 4:00:00 AM
Categories: cs.LG, cs.AI
arXiv:2607.16169v2 Announce Type: replace Abstract: Muon is competitive with AdamW in large-scale pre-training, but its value for reinforcement-learning (RL) post-training remains unclear. We study vanilla Muon in sparse-reward agentic RL through matched single-seed comparisons with AdamW on ALFWorl...
409. Nonconvex Matrix Factorization is Geodesically Convex: Global Landscape Analysis for Fixed-rank Matrix Optimization From a Riemannian Perspective ​
Author: Yuetian Luo, Nicolas Garcia Trillos
Published: 7/21/2026, 4:00:00 AM
Categories: math.OC, cs.IT, cs.LG, cs.NA, eess.SP, math.IT, math.NA
arXiv:2209.15130v3 Announce Type: replace-cross Abstract: We study a general matrix optimization problem with a fixed-rank positive semidefinite (PSD) constraint. We perform the Burer-Monteiro factorization and consider a particular Riemannian quotient geometry in a search space that has a total spa...
410. Gradient-Free Privacy Leakage in Federated Language Models through Selective Weight Tampering ​
Author: Md Rafi Ur Rashid, Vishnu Asutosh Dasu, Kang Gu, Najrin Sultana, Shagufta Mehnaz
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2310.16152v5 Announce Type: replace-cross Abstract: Federated learning (FL) has become a key component in various language modeling applications such as machine translation, next-word prediction, and medical record analysis. These applications are trained on datasets from many FL participants ...
411. Gradient Span Algorithms Make Predictable Progress in High Dimension ​
Author: Felix Benning, Leif D"oring
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC, math.PR
arXiv:2410.09973v2 Announce Type: replace-cross Abstract: We prove that all 'gradient span algorithms' have asymptotically deterministic behavior on scaled Gaussian random functions as the dimension tends to infinity. This is a functional generalization of similar results for random quadratic functi...
412. A Benchmark of Generative Methods for Zero-Shot Environmental Sound Classification ​
Author: Ysobel Sims, Alexandre Mendes, Stephan Chalup
Published: 7/21/2026, 4:00:00 AM
Categories: cs.SD, cs.LG, eess.AS
arXiv:2412.03771v4 Announce Type: replace-cross Abstract: Zero-shot learning enables models to generalise to unseen classes using semantic information, bridging the gap between training classes and previously unseen test classes. While widely studied in computer vision, its application to environmen...
413. Regret Minimization for Piecewise Linear Rewards: Contracts, Auctions, and Beyond ​
Author: Francesco Bacchiocchi, Matteo Castiglioni, Alberto Marchesi, Nicola Gatti
Published: 7/21/2026, 4:00:00 AM
Categories: cs.GT, cs.LG
arXiv:2503.01701v2 Announce Type: replace-cross Abstract: Most microeconomic models of interest involve optimizing a piecewise linear function. These include contract design in hidden-action principal-agent problems, selling an item in posted-price auctions, and bidding in first-price auctions. When...
414. Comprehend, Divide, and Conquer: Feature Subspace Exploration via Multi-Agent Hierarchical Reinforcement Learning ​
Author: Weiliang Zhang, Xiaohan Huang, Yi Du, Ziyue Qiao, Qingqing Long, Zhen Meng, Yuanchun Zhou, Meng Xiao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2504.17356v3 Announce Type: replace-cross Abstract: Feature selection aims to preprocess the target dataset, find an optimal and most streamlined feature subset, and enhance the downstream machine learning task. Among filter, wrapper, and embedded-based approaches, the reinforcement learning (...
415. STRATA: A Name-and-Geography Race Inference Model for Fair Lending and Housing Equity Applications ​
Author: S. Chalavadi, A. Pastor, T. Leitch
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CY, cs.LG
arXiv:2504.21259v2 Announce Type: replace-cross Abstract: Accurate imputation of race and ethnicity (R&E) is essential for fair lending compliance under ECOA, HMDA, and the Community Reinvestment Act, where up to 15% of mortgage applications carry missing race data and regulated institutions bear re...
416. Enhancing LLMs' Clinical Reasoning with Real-World Data from a Nationwide Sepsis Registry ​
Author: Junu Kim, Chaeeun Shim, Sungjin Park, Su Yeon Lee, Gee Young Suh, Chae-Man Lim, Seong Jin Choi, Song Mi Moon, Kyoung-Ho Song, Eu Suk Kim, Hong Bin Kim, Sejoong Kim, Chami Im, Dong-Wan Kang, Yong Soo Kim, Hee-Joon Bae, Sung Yoon Lim, Han-Gil Jeong, Edward Choi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2505.02722v2 Announce Type: replace-cross Abstract: Although large language models (LLMs) have demonstrated impressive reasoning capabilities across general domains, their effectiveness in real-world clinical practice remains limited. This is likely due to their insufficient exposure to real-w...
417. OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration ​
Author: Shijun Li, Hilaf Hasson, Joydeep Ghosh
Published: 7/21/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.LG
arXiv:2505.11765v5 Announce Type: replace-cross Abstract: Agents powered by advanced large language models (LLMs) have demonstrated impressive capabilities across diverse complex applications. Recently, Multi-Agent Systems (MAS), wherein multiple agents collaborate and communicate with each other, h...
418. Data Balancing Strategies: A Systematic Survey of Resampling and Augmentation Methods ​
Author: Behnam Yousefimehr, Mehdi Ghatee, Javad Fazli, Shervin Ghaffari, Zahra Rafei, Mohammad Amin Seifi, Sajed Tavakoli, Abolfazl Nikahd, Mahdi Razi Gandomani, Alireza Orouji, Ramtin Mahmoudi Kashani, Sarina Heshmati, Negin Sadat Mousavi
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2505.13518v3 Announce Type: replace-cross Abstract: Imbalanced datasets, where one class significantly outnumbers others, remain a persistent challenge in machine learning, often biasing predictions toward the majority class and degrading classifier performance. This paper provides a comprehen...
419. Chameleon: A Multiplier-Free Temporal Convolutional Network Accelerator for End-to-End Few-Shot and Continual Learning from Sequential Data ​
Author: Douwe den Blanken, Charlotte Frenkel
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AR, cs.LG
arXiv:2505.24852v3 Announce Type: replace-cross Abstract: On-device learning at the edge enables low-latency, private personalization with improved long-term robustness and reduced maintenance costs. Yet, achieving scalable, low-power end-to-end on-chip learning, especially from real-world sequentia...
420. Breaking the Block: Preserving Data Continuity to Train Superior SAEs for Instruct Models ​
Author: Jiaming Li, Haoran Ye, Yukun Chen, Xinyue Li, Lei Zhang, Hamid Alinejad-Rokny, Jimmy Chih-Hsien Peng, Min Yang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2506.07691v2 Announce Type: replace-cross Abstract: Sparse Autoencoders (SAEs) are a cornerstone of mechanistic interpretability. Existing training methods inherit the Block Training paradigm from LLM pre-training, which introduces destructive gradient noise in instruct models due to attention...
421. Stop Misusing t-SNE and UMAP for Visual Analytics ​
Author: Hyeon Jeon, Jeongin Park, Sungbok Shin, Jinwook Seo
Published: 7/21/2026, 4:00:00 AM
Categories: cs.HC, cs.LG
arXiv:2506.08725v3 Announce Type: replace-cross Abstract: Misuses of t-SNE and UMAP in visual analytics have become increasingly common. For example, although t-SNE and UMAP projections often do not faithfully reflect the original distances between clusters, practitioners frequently use them to inve...
422. LogicIF: Towards Complex Logic Instruction Following ​
Author: Mian Zhang, Shujian Liu, Sixun Dong, Ming Yin, Yebowen Hu, Xun Wang, Simin Ma, Song Wang, Sathish Reddy Indurthi, Haoyun Deng, Zhiyu Zoey Chen, Kaiqiang Song
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2508.09125v3 Announce Type: replace-cross Abstract: Instruction following has catalyzed the recent era of Large Language Models (LLMs) and is the foundational skill underpinning more advanced capabilities such as reasoning and agentic behaviors. As tasks grow more challenging, the logic struct...
423. ChipChat: Low-Latency Cascaded Conversational Agent in MLX ​
Author: Tatiana Likhomanenko, Richard He Bai, Zijin Gu, Zakaria Aldeneh, Shiladitya Dutta, Luke Carlson, Han Tran, Yizhe Zhang, Ruixiang Zhang, Huangjie Zheng, Navdeep Jaitly
Published: 7/21/2026, 4:00:00 AM
Categories: eess.AS, cs.CL, cs.LG, cs.SD
arXiv:2509.00078v2 Announce Type: replace-cross Abstract: The emergence of large language models (LLMs) has transformed spoken dialog systems, yet the optimal architecture for real-time on-device voice agents remains an open question. While end-to-end approaches promise theoretical advantages, casca...
424. SATQuest: A Verifier for Logical Reasoning Evaluation and Reinforcement Fine-Tuning of LLMs ​
Author: Yanxiao Zhao, Yaqian Li, Zihao Bo, Rinyoichi Takezoe, Haojia Hui, Mo Guang, Lei Ren, Xiaolin Qin, Kaiwen Long
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO
arXiv:2509.00930v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong general reasoning, yet the community lacks controllable, scalable, and verifiable tools to analyze and improve these abilities. We present SATQuest, a verifier that generates diverse SAT-based reaso...
425. Implicit Actor Critic Coupling via a Supervised Learning Framework for RLVR ​
Author: Jiaming Li, Longze Chen, Ze Gong, Yukun Chen, Lu Wang, Wanwei He, Run Luo, Min Yang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2509.02522v3 Announce Type: replace-cross Abstract: Recent advances in Reinforcement Learning with Verifiable Rewards (RLVR) have empowered large language models (LLMs) to tackle challenging reasoning tasks such as mathematics and programming, however existing RLVR methods often suffer from sp...
426. One-shot acceleration of transient PDE solvers via online-learned preconditioners ​
Author: Mikhail Khodak, Min Ki Jung, Brian Wynne, Edmond Chow, Egemen Kolemen
Published: 7/21/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG, cs.NA, math.NA, stat.ML
arXiv:2509.08765v4 Announce Type: replace-cross Abstract: Data-driven acceleration of scientific computing workflows has been a high-profile aim of machine learning (ML) for science, with numerical simulation of transient partial differential equations (PDEs) being one of the main applications. The ...
427. Representation-Aware Distributionally Robust Optimization: A Knowledge Transfer Framework ​
Author: Zitao Wang, Nian Si, Molei Liu
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ME, cs.LG
arXiv:2509.09371v2 Announce Type: replace-cross Abstract: Distributionally robust optimization (DRO) protects statistical learning against distributional shifts by optimizing the worst-case performance over a set of perturbed distributions. However, standard DRO formulations often treat all feature ...
428. STAC: When Innocent Tools Form Dangerous Chains for LLM Agents ​
Author: Jing-Jing Li, Jianfeng He, Chao Shang, Devang Kulshreshtha, Xun Xian, Yi Zhang, Hang Su, Sandesh Swamy, Yanjun Qi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG
arXiv:2509.25624v3 Announce Type: replace-cross Abstract: As LLMs advance into autonomous agents with tool-use capabilities, they introduce security challenges that extend beyond traditional content-based LLM safety concerns. This paper introduces Sequential Tool Attack Chaining (\STAC), a novel mul...
429. Computations and ML for surjective rational maps ​
Author: Ilya Karzhemanov
Published: 7/21/2026, 4:00:00 AM
Categories: math.AG, cs.LG
arXiv:2510.08093v2 Announce Type: replace-cross Abstract: The present note studies \emph{surjective rational endomorphisms} $f: \mathbb{P}^2 \dashrightarrow \mathbb{P}^2$ with \emph{cubic} terms and the indeterminacy locus $I_f \ne \emptyset$. We develop an experimental approach, based on some Pytho...
430. OmniX: From Unified Panoramic Generation and Perception to Graphics-Ready 3D Scenes ​
Author: Yukun Huang, Jiwen Yu, Yanning Zhou, Jianan Wang, Xintao Wang, Pengfei Wan, Xihui Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.LG
arXiv:2510.26800v2 Announce Type: replace-cross Abstract: There are two prevalent ways for automatic 3D scene construction: procedural generation and 2D lifting. Among these, panorama-based 2D lifting has emerged as a promising technique, leveraging powerful 2D generative priors to produce immersive...
431. BUSTR: Descriptor-Aware Vision-Language Learning for Breast Ultrasound Report Generation ​
Author: Rawa Mohammed, Mina Attin, Laxmi Gewali, Bryar Shareef
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG
arXiv:2511.20956v2 Announce Type: replace-cross Abstract: Breast ultrasound (BUS) reporting relies on clinically meaningful lesion descriptors, including BI-RADS category, lesion shape, margin, echogenicity, posterior features, pathology, and histology. However, many public BUS datasets provide stru...
432. Towards AI epidemiology: a measurement standardisation framework for prospective risk detection ​
Author: Kit Tempest-Walters
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2512.15783v4 Announce Type: replace-cross Abstract: This paper proposes a measurement standardisation framework that compresses expert-AI interactions into structured, comparable fields for prospective risk detection in deployed AI systems, without access to model internals. This concept paper...
433. Lil: Less is Less When Applying Post-Training Sparse-Attention Algorithms in Long-Decode Stage ​
Author: Junhao Hu, Fangze Li, Mingtao Xu, Feifan Meng, Shiju Zhao, Tiancheng Hu, Ting Peng, Anmin Liu, Wenrui Huang, Chenxu Liu, Ziyue Hua, Tao Xie
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2601.03043v4 Announce Type: replace-cross Abstract: Large language models (LLMs) demonstrate strong capabilities across a wide range of complex tasks and are increasingly deployed at scale, placing significant demands on inference efficiency. Prior work typically decomposes inference into pref...
434. Sockpuppetting: Jailbreaking LLMs by Combining Prefilling with Optimization ​
Author: Asen Dotsinski, Panagiotis Eustratiadis
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.CR, cs.LG
arXiv:2601.13359v3 Announce Type: replace-cross Abstract: Prefill attacks are an effective and low-cost jailbreaking method, as they directly insert an acceptance sequence (e.g., "Sure, here is...") at the start of an LLM's output and lead the model to continue the response. We make two contribution...
435. Solving the Offline and Online Min-Max Problem of Non-smooth Submodular-Concave Functions: A Zeroth-Order Approach ​
Author: Amir Ali Farzin, Yuen-Man Pun, Philipp Braun, Tyler Summers, Iman Shames
Published: 7/21/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.NA, math.NA
arXiv:2601.21243v3 Announce Type: replace-cross Abstract: We consider max-min and min-max problems with objective functions that are possibly non-smooth, submodular with respect to the minimiser and concave with respect to the maximiser. We investigate the performance of a zeroth-order method applie...
436. Singular Bayesian Neural Networks ​
Author: Mame Diarra Toure, David A. Stephens
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.AP
arXiv:2602.00387v4 Announce Type: replace-cross Abstract: Bayesian neural networks promise calibrated uncertainty but require $O(mn)$ parameters for standard mean-field Gaussian posteriors. We argue this cost is often unnecessary, particularly when weight matrices exhibit fast singular value decay. ...
437. Hippasus: Effective and Efficient Automatic Feature Augmentation for Machine Learning Tasks on Relational Data ​
Author: Serafeim Papadias, Kostas Patroumpas, Dimitrios Skoutas
Published: 7/21/2026, 4:00:00 AM
Categories: cs.DB, cs.LG
arXiv:2602.02025v2 Announce Type: replace-cross Abstract: ML models critically depend on feature quality, yet in real-world settings, useful features are often distributed across multiple relational tables rather than a single dataset. Feature augmentation addresses this problem by automatically dis...
438. Uncertainty Decomposition for Bayes-Filtered Transformers via Bayesian Predictive Inference ​
Author: Sandra Fortini, Kenyon Ng, Sonia Petrone, Judith Rousseau, Susan Wei
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2602.04596v2 Announce Type: replace-cross Abstract: Bayes-filtered transformers are transformers meta-learned on sequences from a prior predictive distribution to approximate the corresponding posterior predictive distribution. They output total predictive uncertainty in a single forward pass ...
439. Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making ​
Author: Ruoyu Chen, Shangquan Sun, Xiaoqing Guo, Sanyi Zhang, Kangwei Liu, Shiming Liu, Zhangcheng Wang, Qunli Zhang, Wei Wang, Hua Zhang, Xiaochun Cao
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2602.07008v4 Announce Type: replace-cross Abstract: Reliable models should not only predict correctly, but also justify decisions with acceptable evidence. Yet conventional supervised learning typically provides only class-level labels, allowing models to achieve high accuracy through shortcut...
440. Bayesian Signal Component Decomposition via Diffusion-within-Gibbs Sampling ​
Author: Yi Zhang, Rui Guo, Yonina C. Eldar
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SP, cs.LG
arXiv:2602.10792v2 Announce Type: replace-cross Abstract: In signal processing, the data collected from sensing devices is often a noisy linear superposition of multiple components, and the estimation of components of interest constitutes a crucial pre-processing step. In this work, we develop a Bay...
441. A Unified Physics-Informed Neural Network for Modeling Coupled Electro- and Elastodynamic Wave Propagation Using Three-Stage Loss Optimization ​
Author: Suhas Suresh Bharadwaj, Reuben Thomas Thovelil
Published: 7/21/2026, 4:00:00 AM
Categories: cs.NE, cs.LG, physics.comp-ph
arXiv:2602.13811v2 Announce Type: replace-cross Abstract: Physics-Informed Neural Networks present a novel approach in SciML that integrates physical laws in the form of partial differential equations directly into the NN through soft constraints in the loss function. This work studies the applicati...
442. Dynamic Decision-Making under Model Misspecification: A Stochastic Stability Approach ​
Author: Xinyu Dai, Daniel Chen, Yian Qian
Published: 7/21/2026, 4:00:00 AM
Categories: econ.TH, cs.LG, math.ST, stat.TH
arXiv:2602.17086v2 Announce Type: replace-cross Abstract: Dynamic decision-making under model uncertainty is central to many economic environments, yet existing bandit and reinforcement learning algorithms rely on the assumption of correct model specification. This paper studies the behavior and per...
443. Exploiting Low-Rank Objective Structure in Discrete Quadratic Optimization ​
Author: Ria Stevens, Fangshuo Liao, Barbara Su, Thanasis Hadjidimoulas, Jianqiang Li, Anastasios Kyrillidis
Published: 7/21/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, math.OC, quant-ph
arXiv:2602.20376v3 Announce Type: replace-cross Abstract: We study the problem of maximizing a complex-valued quadratic form over the $K^{\text{th}}$ roots of unity. We show that when the objective matrix $\mathbf{Q}^\star \in \mathbb{C}^{n \times n}$ of the quadratic has rank $r$, the global maximi...
444. No Certificate, No Categorical Speech Act: A Brouwerian Assertibility Constraint for Public Reason ​
Author: Michael J"ulich
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.LG, cs.LO
arXiv:2603.03971v4 Announce Type: replace-cross Abstract: Generative AI can convert uncertainty into authoritative-seeming verdicts, displacing the justificatory work on which democratic epistemic agency depends. As a corrective, I propose a Brouwer-inspired assertibility constraint for responsible ...
445. SCA: Segment-Wise CoT Compression with Answer Alignment ​
Author: Ye Tian, Hongyu Lin
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2603.07598v2 Announce Type: replace-cross Abstract: Chain-of-thought (CoT) reasoning improves problem solving, but long think traces increase inference cost. Existing CoT compression methods usually optimize completion-level length. For structured thinking models, however, a completion contain...
446. The Truncation Blind Spot: How Decoding Strategies Systematically Exclude Human-Like Token Choices ​
Author: Esteban Garces Arias, Nurzhan Sapargali, Christian Heumann, Matthias A{\ss}enmacher
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, stat.ML
arXiv:2603.18482v3 Announce Type: replace-cross Abstract: Why does machine-generated text remain detectable? We trace the answer to the decoding stage: standard strategies such as top-$k$ and nucleus sampling restrict generation to high-probability tokens, while human writers routinely choose words ...
447. HiCI: Hierarchical Construction-Integration for Long-Context Attention ​
Author: Xiangyu Zeng, Qi Xu, Yunke Wang, Chang Xu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2603.20843v3 Announce Type: replace-cross Abstract: Long-context language modeling is commonly framed as a scalability challenge of token-level attention, yet local-to-global information structuring remains largely implicit in existing approaches. Drawing on cognitive theories of discourse com...
448. Spectral-Transport Stability and Benign Overfitting in Interpolating Learning ​
Author: Gustav Olaf Yunus Laitinen-Lundstr"om Fredriksson-Imanov
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2604.08625v2 Announce Type: replace-cross Abstract: We develop a theoretical framework for generalization in the interpolating regime of statistical learning. The central question is why highly overparameterized estimators can attain zero empirical risk while still achieving nontrivial predict...
449. MinShap: A Shapley-Based Framework for Feature Redundancy ​
Author: Chenghui Zheng, Garvesh Raskutti
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2604.15107v2 Announce Type: replace-cross Abstract: Shapley values provide a flexible framework for attributing feature contributions to model predictions, but they are not naturally suited for feature selection: a feature may receive a positive attribution even when it is redundant given the ...
450. Turtle shell clustering: A mixture approach to discriminative clustering with applications to flow cytometry and other data ​
Author: Mackenzie R. Neal, Paul D. McNicholas, Arthur White
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME
arXiv:2604.23083v2 Announce Type: replace-cross Abstract: Generative approaches to clustering provide information on geometric properties of clusters, whereas discriminative approaches provide boundaries between clusters. Ideas from both approaches are incorporated to present a fully unsupervised, p...
451. Information-Theoretic Measures in AI: A Practical Decision Framework ​
Author: Nikolaos Al. Papadopoulos, Konstantinos E. Psannis
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.IT, cs.LG, cs.MA, math.IT
arXiv:2604.23716v3 Announce Type: replace-cross Abstract: Information-theoretic (IT) measures are ubiquitous in artificial intelligence: entropy drives decision-tree splits and uncertainty quantification, cross-entropy is the default classification loss, mutual information underpins representation l...
452. Perturbation is All You Need for Extrapolating Language Models ​
Author: Zetai Cen, Jin Zhu, Xinwei Shen, Chengchun Shi
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH
arXiv:2605.04344v2 Announce Type: replace-cross Abstract: This paper develops a statistical theory of extrapolation for large language models, by reinterpreting them through pre-post-additive noise models. In contrast to the standard autoregressive next-token prediction based on an exact prefix, we ...
453. NaP-Control: Navigating Diffusion Prior for Versatile and Fast Character Control ​
Author: Chia-Wen Chen, Yan Wu, Korrawe Karunratanakul, Siyu Tang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.GR, cs.LG, cs.RO
arXiv:2605.20209v2 Announce Type: replace-cross Abstract: Achieving precise, versatile whole-body character control in physics-based animation remains challenging. Recent diffusion-based policies generate rich and expressive motions but typically rely on gradient-based test-time guidance to satisfy ...
454. FAME: Failure-Aware Mixture-of-Experts for Message-Level Log Anomaly Detection ​
Author: Huanchi Wang, Zihang Huang, Yifang Tian, Kristina Dzeparoska, Hans-Arno Jacobsen, Alberto Leon-Garcia
Published: 7/21/2026, 4:00:00 AM
Categories: cs.SE, cs.LG
arXiv:2605.22779v2 Announce Type: replace-cross Abstract: Production systems generate millions of log lines daily, yet most anomaly detectors operate at the session or window-level, flagging groups of lines rather than identifying the specific message responsible. This coarse granularity forces oper...
455. WINO: A Weak-Form Physics Informed Neural Operator for Hyperelasticity on Variable Domains ​
Author: Bokai Zhu, Yizheng Wang, Qinghui Zhang, Timon Rabczuk
Published: 7/21/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA
arXiv:2605.24651v2 Announce Type: replace-cross Abstract: We propose a Weak-form Physics-Informed Neural Operator (WINO), a data-free framework that combines the efficiency of neural operators with the geometric flexibility of the $\varphi$-finite element method ($\varphi$-FEM). $\varphi$-FEM is an ...
456. Geometry Adaptive Counterfactual Distribution Learning with Diffusion-Guided Smoothing ​
Author: Kwangho Kim
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML
arXiv:2605.25811v2 Announce Type: replace-cross Abstract: We study counterfactual distribution learning for high-dimensional outcomes whose laws may concentrate near lower-dimensional structure. Standard isotropic smoothing ignores this geometry, leading to unfavorable scaling and unstable local inf...
457. SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? ​
Author: Hwiwon Lee, Jiawei Liu, Dongjun Kim, Wubing Xia, Ziqi Zhang, Chunqiu Steven Xia, Lingming Zhang
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CR, cs.LG
arXiv:2605.26548v2 Announce Type: replace-cross Abstract: Finding a real vulnerability in complicated systems is a challenging, long-horizon task that demands reasoning across an entire codebase to produce a working proof-of-concept (PoC). However, such critical security problems remain understudied...
458. Do We Really Need Quantum Machine Learning?: A Multidimensional Empirical Study ​
Author: Sudip Vhaduri, Ryan Gammon, Sayanton Dibbo
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, quant-ph
arXiv:2605.27923v2 Announce Type: replace-cross Abstract: The rapid growth of computer vision and increasingly complex image recognition tasks has exposed fundamental computational limitations of classical machine learning models, motivating the exploration of quantum computing as an emerging new pa...
459. Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief ​
Author: Hongqiang Lin, Pengfei Wang, Nenggan Zheng
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2606.00680v3 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) aims to optimize policies from pre-collected datasets. A bottleneck of this paradigm is managing epistemic uncertainty, which arises from limited data coverage (sample-level) and the ambiguity in identifyin...
460. Auditing Asset-Specific Preferences in Financial Large Language Models: Evidence from Bitcoin Representations and Portfolio Allocation ​
Author: Wenbin Wu
Published: 7/21/2026, 4:00:00 AM
Categories: q-fin.GN, cs.CY, cs.LG
arXiv:2606.02528v3 Announce Type: replace-cross Abstract: Large language models now power robo-advisors and trading agents, yet whether they carry built-in biases toward specific assets is largely untested. We ask three questions: do LLMs systematically prefer certain financial instruments; can an i...
461. CaloTrilogy: Toward a Breakthrough in One-Step, End-to-End, Physics-Guided Shower Generation for Modern Calorimeters ​
Author: Cheng Jiang, Sitian Qian, Kevin Pedro, Oz Amram, Huilin Qu, Maggie Voetberg
Published: 7/21/2026, 4:00:00 AM
Categories: hep-ex, cs.LG, hep-ph, physics.ins-det
arXiv:2606.04165v2 Announce Type: replace-cross Abstract: High-precision calorimeter simulation at current and future colliders imposes rapidly growing computational demands, motivating the development of machine-learning surrogates for traditional Monte Carlo tools such as Geant4. Flow matching and...
462. CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities ​
Author: Tianneng Shi, Robin Rheem, Dongwei Jiang, Mona Wang, Francisco De La Riega, Zhun Wang, Jingzhi Jiang, Alexander Cheung, Sean Tai, Jonah Cha, Jianhong Tu, Gabriel Han, Chenguang Wang, Jingxuan He, Wenbo Guo, Dawn Song
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG
arXiv:2606.04460v2 Announce Type: replace-cross Abstract: AI has the potential to transform cybersecurity by enabling systems that can autonomously detect, analyze, and remediate software vulnerabilities. However, existing cybersecurity evaluations of AI systems are limited in scale or scope, and fa...
463. Physics-Guided Spatiotemporal Learning for Coastal Wave Peak Period Estimation from Video ​
Author: Abubakar Hamisu Kamagata, Dharm Singh Jat, Attlee Munyaradzi Gamundani, Abhishek Srivastava, Paramasivam Saravanakumar
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2606.13302v2 Announce Type: replace-cross Abstract: Direct estimation of physically interpretable periodic signals from raw video constitutes a spatiotemporally grounded learning problem that proves to be difficult especially when facing label sparsity, lack of physical grounding and standardi...
464. Spectral Adaptive Conformal Prediction for Structured Non-Exchangeable Data ​
Author: Jeffery Opoku, David Banahene
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2606.15950v2 Announce Type: replace-cross Abstract: Conformal prediction gives prediction intervals with finite-sample coverage when the data are exchangeable. Many time-indexed datasets are not exchangeable: they have seasons, recurring regimes, changing frequencies, or other forms of structu...
465. Reclaim Evaluation: A Lossy Memory Is Worse Than an Empty One ​
Author: Alex Kwon
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2606.25449v5 Announce Type: replace-cross Abstract: A language model's memory can be worse than no memory at all when the model or its interface is disposed to act on it: a memory that keeps a wrong conclusion but drops the work behind it leads a model to re-emit the stale value as a confident...
466. Learning to Fold: prizewinning solution at LeHome Challenge 2026 (1st place online, 2nd offline) ​
Author: Ilia Larchenko
Published: 7/21/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG
arXiv:2606.27163v2 Announce Type: replace-cross Abstract: I describe my solution to the LeHome Challenge 2026, an ICRA 2026 competition on bimanual garment folding. The system placed 1st of 62 teams in the online (simulation) round and 2nd in the real-world final. It improves a vision-language-actio...
467. Theoria: Rewrite-Acceptability Verification over Informal Reasoning States ​
Author: Michael Saldivar, Ben Slivinski
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.LO, cs.SE
arXiv:2607.01223v4 Announce Type: replace-cross Abstract: When should an AI system's answer be trusted? Formal proof assistants offer certainty but cannot reach most of the problem distribution; scalar LLM judges offer coverage but produce opaque scores that cannot be audited after the fact and are ...
468. Full Bayesian Reinforcement Learning via LF-IBIS ​
Author: Stefano Masini, Cecilia Viscardi, Michela Baccini
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG
arXiv:2607.01741v2 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) is a sequential decision-making framework in which an agent learns optimal policies through interaction with an environment by maximizing cumulative rewards. Among RL methods, Bayesian Reinforcement Learning (BRL) ...
469. Towards transferable lightweight neuromorphic computing through a model-free temporal-switch framework ​
Author: Zefeng Zhang, Chao Li, Siyao Chen, Pei Chen, Bo-Wei Qin, Xumeng Zhang, Wei Lin, Qi Liu
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AR, cs.LG, cs.NE
arXiv:2607.02608v2 Announce Type: replace-cross Abstract: Lightweight neuromorphic computing offers a promising route to efficient AI, with particular benefits for resource-constrained edge deployments. However, its scalable deployment that can reliably transfer the expected performance has long bee...
470. Beyond Multilingual Averages: MTEB-PT, a Benchmark for Portuguese Sentence Encoders ​
Author: Lucas Hideki Takeuchi Okamura, Alexandre Alcoforado, Anna Helena Reali Costa
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.04071v2 Announce Type: replace-cross Abstract: Portuguese remains underrepresented in text embedding evaluation, despite being one of the most widely spoken languages in the world. As a result, embedding models are often selected based on English or multilingual metrics, while their effec...
471. EHR-MPC: Inference-Time Control for Sepsis Treatment with Generative Patient Digital Twins ​
Author: Joshua Pickard, Wei Qi, Na Li, Ann Woolley, Lisa Cosimi, Roy Kishony, Deborah Hung
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, cs.SY, eess.SY, math.OC
arXiv:2607.08793v3 Announce Type: replace-cross Abstract: Sepsis is a leading cause of mortality, yet optimal treatment policies remain contested. Existing reinforcement learning (RL) approaches learn fixed strategies for sepsis treatment, limiting adaptability to changing clinical objectives during...
472. Learning-enabled Parameter Synthesis for Nonlinear Systems from Signal Temporal Logic ​
Author: Alex Beaudin, Hanna Krasowski, Eric Palanques-Tost, Calin Belta, Murat Arcak
Published: 7/21/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY
arXiv:2607.08899v2 Announce Type: replace-cross Abstract: Signal Temporal Logic (STL) is increasingly used to describe interpretable objectives and constraints for optimal control and learning methods, especially when no target time series data is available. In this work, we propose to synthesize pa...
473. A small language model detects behavioural faithfulness gaps that frontier judges and human raters miss ​
Author: Kwan Soo Shin, In Seok Kang, Yunkyung Min, Munho Lee
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.HC, cs.LG
arXiv:2607.09306v2 Announce Type: replace-cross Abstract: Whether a language model behaves as it claims is a judgement on which independent human raters cannot agree (Fleiss kappa = 0.074). We show that a small, purpose-built instrument does better. A linear read-out of the frozen representation of ...
474. Did We Actually Fix It? An Independent Adversarial Stress-Test of Post-Point-Adjustment Evaluation Metrics for Time-Series Anomaly Detection ​
Author: Zongye Lyu
Published: 7/21/2026, 4:00:00 AM
Categories: stat.ML, cs.LG
arXiv:2607.11969v2 Announce Type: replace-cross Abstract: Point-adjustment (PA), for years the default scoring protocol in time-series anomaly detection (TSAD), was shown by Kim et al. (2022) to award near-perfect F1 to random anomaly scores. The field adopted a suite of replacement metrics (PA%K, r...
475. Label-Decoupled Style Augmentation for Domain Generalization in Multi-Label Remote Sensing Scene Classification ​
Author: Alaa Almouradi, Erchan Aptoula
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.12704v3 Announce Type: replace-cross Abstract: Multi-label classification assigns several co-occurring labels to each aerial scene, yet deployed models often encounter data distributions different from their training. Feature-statistics augmentation such as MixStyle, EFDMix, and correlate...
476. Mixed-Timescale Differential Coding for Downlink Model Broadcast in Wireless Federated Learning ​
Author: Chung-Hsuan Hu, Zheng Chen, Erik G. Larsson
Published: 7/21/2026, 4:00:00 AM
Categories: cs.IT, cs.DC, cs.LG, eess.SP, math.IT
arXiv:2607.13119v2 Announce Type: replace-cross Abstract: In standard federated learning systems, the parameter server broadcasts the global model to the participating devices in every iteration. Motivated by the temporal correlation between consecutive global models, differential coding can be appl...
477. When a Verified World Model Still Loses: Play-Adequacy vs Prediction-Accuracy in LLM-Synthesized Code World Models ​
Author: Javier Aguilar Mart'in
Published: 7/21/2026, 4:00:00 AM
Categories: cs.AI, cs.LG
arXiv:2607.14169v2 Announce Type: replace-cross Abstract: Large language models can synthesize a game's rules as executable code - a Code World Model (CWM) - which a classical planner then searches over. Such models are typically accepted when they reach high transition accuracy on sampled trajector...
478. MamaBench: Benchmarking LLM Robustness in Maternal and Child Health Diagnosis through Counterfactual Clinical Perturbation ​
Author: Thanni Adewuyi, Anuoluwa Sotome, Samuel Okoko, Angel Ezendu, Oluwafunke Akinbuwa, Oluwaseun Odunsi, Oluwasegun Oguntuase, Ifeoma Nwabueze, Abiodun Adereni
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.LG
arXiv:2607.14385v2 Announce Type: replace-cross Abstract: Large language models achieve strong scores on medical benchmarks, yet these benchmarks evaluate each question in isolation, providing no measure of whether a system can distinguish clinically similar presentations requiring different interve...
479. Decision Making Needs Uncertainty Quantification [Lecture Notes] ​
Author: Osvaldo Simeone
Published: 7/21/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, math.IT
arXiv:2607.14407v2 Announce Type: replace-cross Abstract: Many signal processing systems ultimately exist to {act}. Whenever the state variable that determines the action to be taken by a decision maker, or agent, is uncertain, the way that uncertainty is represented decides how well the agent perfo...
480. Unsupervised Keypoints for Real-Time Fall Detection: Comparative Analysis Under Real-world Conditions with Predictive Bandwidth Reduction ​
Author: Tasmiah Haque, Jacob Kosinski, Sumit Mohan, Mohammad Abdullah Al-Mamun, Srinjoy Das
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CV, cs.LG
arXiv:2607.15400v2 Announce Type: replace-cross Abstract: Falls among older adults are a major safety challenge, but continuous monitoring is difficult to sustain. Video captures fall-related posture and motion, yet deployment is limited by privacy, computation, and bandwidth. Supervised pose estima...
481. Induction in Both Directions: A Mechanistic Analysis of In-Context Learning in Masked Diffusion Language Models ​
Author: Andy Catruna, Emilian Radoi
Published: 7/21/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG
arXiv:2607.15893v2 Announce Type: replace-cross Abstract: While the internal mechanisms of autoregressive (AR) transformers have been studied extensively, much less is known about diffusion language models (DLMs), an emerging alternative that generates text by iterative denoising. In this work, we s...