Skip to content

arXiv cs.LG - 2026-07-24 ​

263 items collected.


1. DataPrep-Bench: Benchmarking LLMs as Training Data Preparators ​

Author: Hao Liang, Qifeng Cai, Yibo Lin, Jianzhuo Du, Qifeng Xia, Sizhe Qiu, Linzhuang Sun, Meiyi Qiang, Zhaoyang Han, Xiaochen Ma, Bohan Zeng, Ruichuan An, Conghui He, Wentao Zhang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.20465v1 Announce Type: new Abstract: The quality of training data fundamentally determines the capabilities of large language models (LLMs), yet no unified benchmark exists to measure how well LLMs, agents, and data-centric workflows actually prepare training data end to end. We view LLM-...

📖 Read original article


2. PhantomFill: When the Form Demands an Answer, Language Models Invent One ​

Author: Rana Muhammad Usman
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.20492v1 Announce Type: new Abstract: Language models in production do not write prose. They fill forms: JSON fields, function arguments, extraction templates. We show that the form itself causes hallucination. We ask thirteen models the same question about the same input and change only t...

📖 Read original article


3. The Active Ingredient in Muon's Grokking ​

Author: Yufeng Wang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20512v1 Announce Type: new Abstract: The Muon optimizer reaches the grokking threshold on modular arithmetic faster than AdamW. Prior work attributes this to "spectral-norm constraints plus orthogonalized momentum" but does not isolate which mechanism matters. To better understand Moun's ...

📖 Read original article


4. Scaling Closed-Loop Feature Channel Configuration with LLMs ​

Author: Tolgay Atinc Uzun, Radu Timofte, Dmitry Ignatov
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20516v1 Announce Type: new Abstract: Promising initial results in closed-loop large-language-model-based channel-configuration search demonstrated that neural-network widths can be optimized directly through executable code generation and accuracy feedback. However, those results were obt...

📖 Read original article


5. Multimodal CoLRAG-TF: Triple-Filtered Retrieval for Complex PDFs ​

Author: Takato Yasuno
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20517v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) over heterogeneous PDF collections remains challenging due to multimodal content, domain-specific terminology, and the need for multi-hop reasoning across dispersed evidence. We present Multimodal CoLRAG-TF, a four-...

📖 Read original article


6. Adaptive Depth in Looped Transformers: Diagnosing Learned Halting Gates and Trajectory Readouts ​

Author: Andrei Cristian Popescu, Haitz S'aez de Oc'ariz Borde, Pietro Li`o
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20519v1 Announce Type: new Abstract: Looped Transformers increase test-time computation by repeatedly applying a shared recurrent block. Learned halting objectives in looped Transformers typically use a single exit distribution both as the inference-time stopping rule and as the training-...

📖 Read original article


7. Generative Bayesian Filtering for State Estimation ​

Author: Lei Cao, Sihang Feng, Jixin Yan, Tao Sun, Naichen Shi
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.20521v1 Announce Type: new Abstract: The state of a dynamic system evolves over time, switching among several latent modes that govern its observable behavior. Filtering methods infer the latent state from observations. Classical filtering approaches, including Kalman filters, typically r...

📖 Read original article


8. Do Active SAE Feature Planes Carry More Holonomy? A Preregistered Reversal in Gemma ​

Author: Larry Richards
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20522v1 Announce Type: new Abstract: This paper tests whether holonomy concentrates on active sparse-autoencoder (SAE) feature planes in Gemma 2 2B, a concrete operationalization of the broader semantic-concentration prediction. Holonomy is measured at the final-token layer-12 to layer-13...

📖 Read original article


9. Uncertainty-Aware Trust Estimation for Multi-LLM Systems via Structured Expert Judgement ​

Author: Jiawei Zheng, Jiazhen Zhang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20529v1 Announce Type: new Abstract: Large Language Model (LLM) ensembles are increasingly used to improve reliability by combining predictions from multiple LLMs. However, existing aggregation methods typically assume that all models are equally trustworthy, overlooking differences in un...

📖 Read original article


10. CLOE: Christoffel Loss Autoencoder for Anomaly Detection ​

Author: L'ea Billet (LAAS, INSA Toulouse, ANITI), Louise Trav'e-Massuy`es (LAAS-DISCO, Comue de Toulouse, ANITI), Elodie Chanthery (LAAS), Alexandre Gaffet
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.20530v1 Announce Type: new Abstract: Semi-supervised anomaly detection plays a key role in diverse fields such as process monitoring, healthcare, and finance. However, lightweight methods often struggle with high-dimensional data and typically require careful tuning of multiple hyperparam...

📖 Read original article


11. Position: Stop Reactively Patching Your Model Every Time and Start Proactive Test-Driven AI Development ​

Author: Nadine Chang, Maying Shen, Jialiang Wang, Rafid Mahmood, Jose M. Alvarez
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20532v1 Announce Type: new Abstract: Many modern AI systems are designed to operate under diverse, open-ended, use-cases. To help generalize deployed systems, many deployed-system maintenance pipelines use a reactive AI flywheel that observes emerging feedback from user behavior (errors) ...

📖 Read original article


12. Grounding Investor Views: Neural Predicates in the Black-Litterman Model ​

Author: Marcos Florencio
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20533v1 Announce Type: new Abstract: Portfolio construction under the Black-Litterman model requires investors to specify views on asset returns alongside explicit uncertainty estimates -- a process that remains largely subjective and difficult to scale. We propose a formal approach in wh...

📖 Read original article


13. A Graph Neural Network approach to zero-shot Digital Twins ​

Author: Alicia Tierz, Ic'iar Alfaro, David Gonz'alez, El'ias Cueto
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.20535v1 Announce Type: new Abstract: Traditional Predictive Digital Twins often remain geometrically rigid, requiring extensive retraining or fine-tuning whenever the underlying physical domain or boundary conditions change. To overcome this limitation, we present a novel framework for \t...

📖 Read original article


14. ReliableTableQA:How Much Supervision Does Reliability Annotation Need? ​

Author: Huei-Chung Hu, Hsin-Tai Wu, Koyo Kobayashi
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20537v1 Announce Type: new Abstract: We introduce ReliableTableQA, a framework for training an LLM to annotate the statistical reliability of tabular QA results, not whether the query is answerable, but whether the computed answer is statistically meaningful. In real enterprise analytics,...

📖 Read original article


15. Codec-Gauge: Learning Compression-Friendly Gauges for Transformer KV Caches ​

Author: Yitao Jiang, Yaoqing Yang, Luyang Zhao, Muhao Chen, Devin Balkcom
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20538v1 Announce Type: new Abstract: Long-context Transformer inference increasingly relies on KV-cache compression or quantization. Prior rotation and transform-coding results suggest that the channel basis of each key/value vector affects how faithfully a fixed backend preserves model b...

📖 Read original article


16. Leveraging Biokinetic Knowledge Priors for Data-Scarce Bioprocess Modeling ​

Author: Kyunghoon Hur, Eunjung Jeon, Hyun Woo Kim, Gyubok Lee, Seongjun Yang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20539v1 Announce Type: new Abstract: While deep learning has accelerated drug discovery, its impact on biomanufacturing has been considerably more limited. The reason is data scarcity. Bioreactor experiments are high-cost, take days to weeks, and are rarely shared in public form, leaving ...

📖 Read original article


17. From Atoms to Entropy: Optimal Noise Allocation for Diffusion Training in the Convex Regime ​

Author: Luca Ambrogioni, Giulio Franzese, Alberto Foresti, Gabriel Raya, Bac Nguyen, Georgios Batzolis, Yuhta Takida, Naoki Murata, Chieh-Hsin Lai, Yuki Mitsufuji
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20540v1 Announce Type: new Abstract: How should a diffusion model decide which noise levels to train on, and how much? Despite the importance of this choice, current noise schedules are based largely on heuristics or empirical tuning. Here, we develop a general statistical framework for s...

📖 Read original article


18. HypNO: A Graph-Based Neural Operator with Physics-Informed Message Passing for Hyperbolic Conservation Laws ​

Author: Dimitrije \v{Z}drale, Cassie An Jeng, Katie Wang, Sonia Vanier, Alexandre Bayen, Hossein Nick Zinat Matin
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20541v1 Announce Type: new Abstract: We introduce HypNO, a graph-based neural operator for scalar hyperbolic conservation laws. HypNO operates directly on a space-time graph of finite-volume cells and uses adjacency-factored, physics-informed message passing to respect upwinding and entro...

📖 Read original article


19. Improving Access to Essential Medicines via Decision-Aware Machine Learning ​

Author: Angel Tsai-Hsuan Chung, Jatu Abdulai, Patrick Bayoh, Lawrence Sandi, Francis Smart, Hamsa Bastani, Osbert Bastani
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY

arXiv:2607.20542v1 Announce Type: new Abstract: A critical challenge in healthcare systems in low- and middle-income countries (LMICs) is the efficient and equitable allocation of scarce resources, particularly essential medicines. This problem is complicated by limited high-quality data, which rest...

📖 Read original article


20. When RLVR Shrinks the Reasoning Boundary: Diagnosing Pass@k Inversion ​

Author: Todd Zhou
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20543v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve one-sample accuracy while making a model worse under repeated sampling. We study this pass@k inversion: after training, the policy may solve fewer distinct problems than its base model a...

📖 Read original article


21. Conflict Resolution under Degraded Surveillance in Air Corridors Using Multi-Agent Reinforcement Learning ​

Author: Esrat Farhana Dulia, Syed Arbab Mohd Shihab, Caleb Adams, Ruben Del Rosario
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20547v1 Announce Type: new Abstract: Safe Advanced Air Mobility operations require aircraft to maintain separation when surveillance information is noisy, delayed, incomplete, or temporarily unavailable. This study develops a Deep Q-Network-based Multi-Agent Reinforcement Learning framewo...

📖 Read original article


22. SOAP, Muon, and Beyond: Pushing LLM Pretraining Scales ​

Author: Mikail Khona, Aditya Vavre, Boxiang Wang, Deyu Fu, Hao Wu, Mike Chrzanowski, Bryan Catanzaro, Dheevatsa Mudigere, Jeff Pool, Michael Lightstone, Mohammad Shoeybi, Mostofa Patwary, Nima Tajbakhsh, Tijmen Blankevoort
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20548v1 Announce Type: new Abstract: Higher-order optimizers such as Muon and SOAP offer faster convergence than AdamW, but their computational cost and numerical stability challenges have limited adoption at scale. In this work, we adapt and enhance preconditioned gradient methods to ove...

📖 Read original article


23. SevDiff: Severity-Conditioned Diffusion for Long-Tail Conflict Trajectory Generation ​

Author: Eni Solomon Laughter
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.RO, cs.SY, eess.SY

arXiv:2607.20549v1 Announce Type: new Abstract: Trajectory datasets used in ADAS evaluation are heavily biased toward routine driving; genuine vehicle-to-vehicle conflict events are rare, and the rarer the event, the higher the cost when an ADAS system fails to handle it. Existing generative approac...

📖 Read original article


24. Beyond SBDD: Geometric Deep Learning in Polypharmacology and Multi-target Drug Design ​

Author: Tianming Han, Zhijie Pan, Wenchi Ge, Qi Zhao
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20550v1 Announce Type: new Abstract: The traditional "one drug, one target" paradigm of structure-based drug design (SBDD) frequently proves inadequate for treating multifactorial diseases such as cancer and neurodegenerative disorders, owing to compensatory signaling pathways and the eme...

📖 Read original article


25. SenCos-GEM: SENet-Calibrated and Law-of-Cosines-Constrained Geometry-Enhanced Molecular Representation for Property Prediction ​

Author: Tianming Han, Li Zhang, Qi Zhao
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, q-bio.BM

arXiv:2607.20551v1 Announce Type: new Abstract: Effective molecular representation learning is crucial for accurate molecular property prediction. Recently, numerous self-supervised learning (SSL) approaches leveraging 3D GNNs have been developed to capture comprehensive 3D structural information fo...

📖 Read original article


26. Thermodynamic Weight Decay: Exploring Grokking Acceleration via Attention Specific Heat ​

Author: Chitraansh Pandey
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20552v1 Announce Type: new Abstract: Grokking -- the delayed generalization of neural networks long after they have memorized their training data -- wastes thousands of training epochs and is notoriously unpredictable. Building on the recent result that Transformer attention is formally i...

📖 Read original article


27. Double-Scoring: Reliable Extraction of Strong Lottery Tickets ​

Author: Bryce A. Christopherson, Jack Baretz, Darian Colgrove, Salah Dandan
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2607.20555v1 Announce Type: new Abstract: The lottery ticket hypothesis proposes that large random neural networks contain sparse subnetworks that can match the performance of dense models after comparable training. A stronger version asserts that sufficiently overparameterized random networks...

📖 Read original article


28. Monkey King Bang: A Unified Scientific Multimodal Foundation Model ​

Author: Hesen Chen, Xinyu Su, Xiaomeng Yang, Yuetan Lin, Zixiong Yang, Junyi An, Fenglei Cao, Yifeng Jiao, Yunqi Zhang, Yuan Cheng, Zhiyu Tan, Hao Li, Libo Wu, Yuan Qi
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20557v1 Announce Type: new Abstract: Scientific discovery is increasingly shifting from isolated disciplines to multi-domain reasoning, and AI for science faces a similar transition. Existing systems are either specialised for individual domains or unify scientific data mainly through tex...

📖 Read original article


29. StabilityBench: Benchmarking Instability in LLMs ​

Author: Emma Kondrup, Zachary Yang, Anne Imouza, Reihaneh Rabbany
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20558v1 Announce Type: new Abstract: AI Assistants are increasingly deployed in high-stakes settings, such as healthcare or government services. Yet their real-world behavior remains poorly understood due to strong context dependence. Current evaluation protocols follow a defense-in-depth...

📖 Read original article


30. Joint Utilization of Geospatial and census proxies for Autoencoder-Assisted Downscaling (JUGAAD) of socioeconomic indicators in India ​

Author: Aditya Dutt, Paul Gader, Aditya Singh
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20559v1 Announce Type: new Abstract: Monitoring poverty and food security indicators is imperative for addressing socioeconomic challenges in developing nations. A limitation is mismatches in scale between data sources: census data provide geographic coverage, while socioeconomic indicato...

📖 Read original article


31. Chronofy: A Temporal-Logical Decay Architecture for Information Validity in Time-Aware Retrieval-Augmented Generation ​

Author: Muntaser Syed, Marius Silaghi, Sheikh Abujar, Sharun Akter
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20560v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems retrieve and integrate external knowledge to ground large language model (LLM) outputs. However, current RAG architectures treat all retrieved facts as equally valid regardless of temporal provenance, leadin...

📖 Read original article


32. CT-Merging: Consensus Directions and Task-Level Scaling for LoRA Adapter Merging ​

Author: Keumseo Ryum, Joonhyuk Kang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.20561v1 Announce Type: new Abstract: LoRA adapters provide an efficient way to specialize a pretrained model for many downstream tasks, but deploying one adapter per task requires adapter storage and task selection at inference time. Model merging addresses this issue by combining indepen...

📖 Read original article


33. AI-Driven Surrogate Models for Predicting Electrode-Scale Discharge Behavior in Lithium-Ion Batteries ​

Author: Mengda Xing (CRIL, UA), Jean-Marie Lagniez (CRIL, UA), Alejandro Franco (LRCS)
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20577v1 Announce Type: new Abstract: Physics-based simulations are essential for understanding the electrode-scale discharge behavior of lithium-ion batteries (LIBs) but suffer from prohibitive computational costs. To address this, we introduce a novel deep learning surrogate pipeline bas...

📖 Read original article


34. Fisher Widths: Local Learning Geometry and Anisotropic Recovery ​

Author: Vu Khac Ky
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH

arXiv:2607.20578v1 Announce Type: new Abstract: We study Gaussian-width complexity on statistical manifolds through a pair of functionals: the primal Fisher width $w_G(T) = w(G^{1/2}T)$, induced by the Fisher metric, and the inverse-Fisher width $w_{G^{-1}}(T) = w(G^{-1/2}T)$, induced by the inverse...

📖 Read original article


35. Bayesian uncertainty estimation improves clinical decision making in medical AI agents ​

Author: Frederik Hauke, Patrick Wienholt, Christiane Kuhl, Dyke Ferber, Jakob Nikolas Kather, Sven Nebelung, Daniel Truhn
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA

arXiv:2607.20582v1 Announce Type: new Abstract: Machine learning models for medical image analysis typically lack a reliable measure of confidence, limiting their use in ambiguous or atypical cases. Here we show that Monte Carlo dropout, applied to a multi-task chest-radiograph classifier (eight tho...

📖 Read original article


36. Exact ReLU realization of affine one-dimensional refinement iterates via residual memory and offset frames ​

Author: Boldsaikhan Bolorkhuu, Tsogtgerel Gantumur
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20586v1 Announce Type: new Abstract: We study vector-valued affine refinement operators of the form [ (W\gamma)(t)=\sum_{j\in\mathbb{Z}} A_j\gamma(Mt-j)+B(t), ] with finitely supported matrix mask and compactly supported continuous piecewise linear input and forcing data. Building on the ...

📖 Read original article


37. Detecting Neural Network Failures through Spectral Analysis of Internal Activations ​

Author: Arunan J
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.20590v1 Announce Type: new Abstract: Neural network misclassifications exhibit characteristic spectral instability in internal activations that is invisible at the output layer. This phenomenon is identified and formalized as Spectral Drift -- the frequency-domain distance between consecu...

📖 Read original article


38. STeMP: Spatio-Temporal Modelling Protocol ​

Author: Jan Linnenbrink, Jakub Nowosad, Marvin Ludwig, Anna Frederike Jablotschkin, Fabian Schumacher, Teja Kattenborn, Hanna Meyer
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20592v1 Announce Type: new Abstract: Spatio-temporal machine-learning modelling is an important tool in environmental research. However, machine-learning models are highly sensitive to both the characteristics of the training data, such as its distribution, and methodological choices, inc...

📖 Read original article


39. When Does Recurrence Become an Algorithm? Convergence Selection in Weight-Tied Looped Transformers ​

Author: Tong Zhang, Junhao Hu, Yun Peng, Tao Xie
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.20594v1 Announce Type: new Abstract: When does a weight-tied looped transformer -- one block applied T times -- implement an actual algorithm? We answer with four findings from controlled populations on group word problems. (1) The budget law: free training installs a linear computation f...

📖 Read original article


40. Are Single-Token Sparse Autoencoder Features Causally Necessary? Layer-Depth and SAE-Family Effects ​

Author: Seonglae Cho, Zekun Wu, Kleyton Da Costa, Rishi Kalra, Ilham Wicaksono, Adriano Koshiyama
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.20596v1 Announce Type: new Abstract: Sparse autoencoder (SAE) features are used to interpret and steer large language models, yet whether a feature's causal role is stable across SAE families remains untested. Single-token features that activate on one vocabulary item provide the diagnost...

📖 Read original article


41. One Round Is All You Need: Analytic Federated Learning for Task-Heterogeneous Multi-Label Medical Image Classification ​

Author: Afsaneh Mahanipour, Hana Khamfroush
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20641v1 Announce Type: new Abstract: Federated learning (FL) enables multiple clinical institutions to collaboratively train a shared disease classifier without centralizing patient data. In practice, however, each institution annotates only the pathologies within its area of expertise, s...

📖 Read original article


42. Scaling Interpretable Transformers with Parity Bottleneck Layers ​

Author: Andrew Mack, Kraig Yuheng Tou, Mark Henry, Zhengxun Wu, Lauren Greenspan
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20652v1 Announce Type: new Abstract: Language models are thought to exhibit the phenomenon of superposition, representing many more features than dimensions in their residual streams. Sparse autoencoders (SAEs) are designed to recover such features post-hoc, but training models that are i...

📖 Read original article


43. SalesLoop: Reinforcement Learning from Performance Feedback for Sales Lead Ranking ​

Author: Chenyu Zhang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR

arXiv:2607.20655v1 Announce Type: new Abstract: Lead ranking in Customer Relationship Management (CRM) systems faces a persistent challenge: models achieving high offline accuracy often underperform in production. We identify three fundamental gaps responsible for this disconnect: offline-online met...

📖 Read original article


44. Adaptive Multi-Horizon Reinforcement Learning ​

Author: Manoosh Samiei, Doina Precup, Paul Masset
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20656v1 Announce Type: new Abstract: Effective decision-making in complex and changing environments requires balancing short-term and long-term consequences. In reinforcement learning (RL), this trade-off is typically controlled through a fixed discount factor, which imposes a single expo...

📖 Read original article


45. End-to-End Learning of Safe Optimal Feedback Control in High Dimensions with Control Barrier Function Layers ​

Author: Xingjian Li, Kelvin Kan, Deepanshu Verma, Krishna Kumar, Stanley Osher, Samy Wu Fung
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY, math.OC

arXiv:2607.20674v1 Announce Type: new Abstract: We consider the problem of learning high-dimensional semi-global feedback controllers under hard safety constraints enforced by control barrier functions (CBFs). Incorporating CBFs into end-to-end policy training requires embedding a quadratic-program-...

📖 Read original article


46. Explanation-Based Runtime Verification for Trustworthy ML-driven Optical Networks ​

Author: Omran Ayoub, Carlos Natalino, Ali Al Housseini, Felix Foschum, Philipp Morger, Tiziano Leidi, David Hock, Paolo Monti
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.NI

arXiv:2607.20675v1 Announce Type: new Abstract: Machine learning (ML) models are increasingly integrated into optical network automation frameworks to support tasks such as failure management, performance monitoring and resource allocation. In these environments, ML-driven predictions may be directl...

📖 Read original article


47. Attribution Markets: A Fisher-Market Formulation for Fractional Credit Assignment Between Planned Tasks and Performed Actions ​

Author: Salavat Ishbulatov
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.GT

arXiv:2607.20694v1 Announce Type: new Abstract: Personal and organizational planning systems maintain two records that drift apart: what was planned (a task's effort budget) and what was done (a logged action's duration and description). Existing systems bridge them with an exclusive, all-or-nothing...

📖 Read original article


48. CEDAR: Causal Edge Discovery for Autoregressive Processes ​

Author: Mohammad Fesanghary
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ME

arXiv:2607.20696v1 Announce Type: new Abstract: We propose CEDAR (Causal Edge Discovery for Autoregressive Processes), a constraint-based method for lagged causal edge discovery in sparse autoregressive time series. CEDAR screens candidate cross-variable lags using AR(1)-residualized, U-centered dis...

📖 Read original article


49. Perspective Latents as an Architectural Condition for Causal Emergence in Active Inference Agents ​

Author: Hongju Pae
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC

arXiv:2607.20708v1 Announce Type: new Abstract: A recent line of work measures causal emergence in reinforcement learning agents through Integrated Information Decomposition, reporting that $\Phi_r$ grows with training and tracks reward improvement. For active inference, this raises the question of ...

📖 Read original article


50. LLMs Get Lost in Evolving User Intent ​

Author: Jihoon Tack, Philippe Laban, Jennifer Neville
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20734v1 Announce Type: new Abstract: As LLMs become more capable, they are increasingly deployed as collaborative agents, taking on user-delegated tasks through iterative interaction. Yet genuine interaction is inherently dynamic: users rarely specify their intent upfront, instead disclos...

📖 Read original article


51. Cardinality-Decomposed Loss: Matching Training Objectives to Relation Structure in Heterogeneous Recommendation Graphs ​

Author: Parul Maheshwari, Amulya Paruchuri, Yiqing Zou, Alireza Sahami Shirazi, Farhad Farahani, Prakhar Mehrotra
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2607.20737v1 Announce Type: new Abstract: Graph Neural Networks trained on heterogenous bipartite graphs form a common basis in recommendation systems. These graphs often express relations that vary in cardinality, for example, user-item preferences are one-to-many and user-attribute features ...

📖 Read original article


52. Adaptive Confidence-weighted Expansion for Trustworthy Multi-Omics Multimodal Fusion ​

Author: Mohammad Raahemi, Ali Sekhavati, Alireza Maleki, Hamid Nasiri
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20742v1 Announce Type: new Abstract: Multimodal learning is a robust approach to improve predictive performance in applications such as medical prognosis. However, the clinical applicability of models that use multimodal learning is hampered by their poor performance under noisy or uninfo...

📖 Read original article


53. GaugeQuant: Online Learning of Quantization-Optimal Bases from LLM Symmetries ​

Author: Miguel P. Bento, Jo~ao Seabra
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.20757v1 Announce Type: new Abstract: Transformers are known to have internal continuous symmetries that leave outputs invariant, while modifying quantization. GaugeQuant leverages this in-training by introducing a LogSumExp term to the loss that breaks the symmetries, thus selecting a bas...

📖 Read original article


54. Memory-Computation Tradeoffs in Semi Amortized Parametric Optimization ​

Author: Shijie Pan, Agustin Castellano, Zeyu Shen, Enrique Mallada
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20769v1 Announce Type: new Abstract: Learning-enabled decision systems often use offline data or computation to reduce online compute cost. Despite the empirical success of such approaches, there is limited general understanding of how much offline information is needed to achieve a desir...

📖 Read original article


55. Toward Mechanistic Interpretability of an AI Foundation Model Fine-Tuned for Atmospheric Chemistry ​

Author: Jason Y. Hu, Ivan Higuera-Mendieta, Patrick Obin Sturm, Makoto M. Kelp
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2607.20778v1 Announce Type: new Abstract: Weather forecasting foundation models (FMs) are increasingly fine-tuned to predict air quality, offering fast global pollution forecasts at lower computational cost than conventional chemical transport models. These FMs are typically trained on reanaly...

📖 Read original article


56. Synthetic minority data is redundant or invalid: a data-dependent validity theory and a de-biased test ​

Author: Ahmad B. Hassanat, Ahmad S. Tarawneh, Ghada A. Altarawneh
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20787v1 Announce Type: new Abstract: For two decades, the standard remedy for class-imbalanced learning has been to fabricate synthetic minority examples, and the standard evidence of their validity has been a check that cannot fail: synthetic points are scored against the very data that ...

📖 Read original article


57. Memoir: Should a Model Write to Its Memory While It Thinks? ​

Author: Jaber Jaber, Osama Jaber
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2607.20792v1 Announce Type: new Abstract: Memoir combines per-sample fast memory, shared slow parameters, variable-depth latent recurrence, and a future-latent energy objective. We test its riskiest coupling: each pondering iteration may rewrite the fast tier that the same iteration reads. On ...

📖 Read original article


58. External Clustering Validation by the Homogeneity-Parsimony Trade-off ​

Author: Andreas Tiffeau-Mayer
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, stat.ME

arXiv:2607.20799v1 Announce Type: new Abstract: Scalar metrics are often used to evaluate clusterings against known classes, but they can obscure a fundamental trade-off: clusterings should be informative about class labels while avoiding unnecessary fragmentation. Here we describe normalized scores...

📖 Read original article


59. New Complexity-Theoretic Frontiers of Tractability for Neural Network Training ​

Author: Cornelius Brand, Robert Ganian, Mathis Rocton
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.DS

arXiv:2607.20811v1 Announce Type: new Abstract: In spite of the fundamental role of neural networks in contemporary machine learning research, our understanding of the computational complexity of optimally training neural networks remains incomplete even when dealing with the simplest kinds of activ...

📖 Read original article


60. Robust Asynchronous Q-Learning under Reward and State Corruption via Batching ​

Author: Sreejeet Maity, Aritra Mitra
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.SY, eess.SY

arXiv:2607.20822v1 Announce Type: new Abstract: Motivated by reinforcement learning in harsh environments, we consider the problem of learning an optimal policy subject to adversarially corrupted feedback. Specifically, at each time-step, an adversary can perturb both the reward and state observatio...

📖 Read original article


61. Beyond Heavy Log Curation: Perplexity-Based APT Detection via Unsupervised, Context-Augmented Language Models ​

Author: Shoya Otsu, Kei Suzuki, Toshiaki Koike-Akino, Jing Liu, Ye Wang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CR

arXiv:2607.20832v1 Announce Type: new Abstract: Advanced Persistent Threats (APTs) remain difficult to detect because only a small fraction of events in large-scale logs are attack-related, and investigation is expensive and hard to scale. Prior machine-learning approaches can reduce analyst workloa...

📖 Read original article


62. Offline RL with Hierarchical Action Chunking ​

Author: Ahad Jawaid
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20834v1 Announce Type: new Abstract: Offline goal-conditioned reinforcement learning (RL) holds the promise of learning general-purpose policies from static datasets. However, scaling these methods to long-horizon tasks remains a challenge due to the curse of horizon, where value estimati...

📖 Read original article


63. Multilevel Graph Wavelet Compressed Sensing with Scale-Aware Neural Recovery ​

Author: Amirhossein Nouranizadeh, Sarang Rajendra Patil, Alan John Varghese, Varsha Narayanan, Amit Chakraborty, Mengjia Xu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20857v1 Announce Type: new Abstract: Scientific machine learning methods such as neural operators and physics-informed neural networks have advanced engineering applications and inverse problems, but their training typically requires large volumes of simulated data. This makes data prepar...

📖 Read original article


64. Position Bias is Hidden Behind Ceiling Effects: A Permutation Diagnostic for LLM Benchmarks ​

Author: Hiroki Tamba
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.20864v1 Announce Type: new Abstract: Position bias in multiple-choice LLM evaluation is widely cited as a confound in capability comparisons, but published measurements rely on single answer-order shuffles whose results confound the bias signal with content-level noise and sampling stocha...

📖 Read original article


65. TwistedMerge: Certified Higher-Order Diagnostics and Abstention for Model Merging ​

Author: Ting Gong, Shitan Xu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.AG

arXiv:2607.20887v1 Announce Type: new Abstract: Model merging combines independently trained or fine-tuned models, but pairwise alignability does not imply globally consistent alignment. We formulate merging as a finite descent problem in which checkpoints are local objects, alignment maps are trans...

📖 Read original article


66. Information-Theoretically Secure Aggregation for Lightweight Federated Learning: Resilient to Dropouts and Adversaries ​

Author: Hyeong-Gun Joo, Songnam Hong, Dong-Joon Shin
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20890v1 Announce Type: new Abstract: On-device federated learning (FL) enables privacy-preserving and personalized model training on resource-constrained devices such as smartphones and IoT nodes. To reduce communication cost, sign-based methods (e.g., signSGD) transmit one-bit gradients....

📖 Read original article


67. HierarchicalDAEW: Domain-Aware Edge-Weighted Graph Convolution with Evidential Uncertainty for Multi-Section Spatial Gene Expression Prediction from H&E Histology ​

Author: Kritanu Chattopadhyay, Soumya Chatterjee, Ondrej Krejcar, Debotosh Bhattacharjee
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, q-bio.GN

arXiv:2607.20896v1 Announce Type: new Abstract: Spatial transcriptomics assays remain costly and technically demanding, restricting transcriptome-wide profiling to specialist settings and preventing routine clinical deployment. Predicting spatially resolved gene expression from H&E histology could c...

📖 Read original article


68. Multi-turn RL with Structural and Performance Aware Rewards for CUDA Kernel Generation ​

Author: Quazi Ishtiaque Mahmud, Nesreen K. Ahmed, Ali Jannesari
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20908v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a powerful technique to enhance the reasoning capacity of LLMs for optimized code generation. However, existing RLVR approaches primarily rely on outcome-based signals such as correct...

📖 Read original article


69. Three-Pronged Spectral Control for Federated Parameter Efficient Fine Tuning ​

Author: Shiva Raj Pokhrel, Dipsan Bhattarai, Anwar Walid
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.NI

arXiv:2607.20914v1 Announce Type: new Abstract: Federated parameter-efficient fine-tuning (PEFT) enables communication-efficient adaptation of large pretrained models on decentralized edge data, but it remains fragile under non-IID client heterogeneity. In low-rank adaptation (LoRA), different clien...

📖 Read original article


70. Best-of-Evidence: Best-of-N Selection under Partial Verification ​

Author: Cenwei Zhang, Teng Fang, Yuxia Wang, Derek Li, Bryan Dai, Lei You
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20950v1 Announce Type: new Abstract: BoN improves model outputs by sampling several candidates and selecting one with a proxy score, but it assumes that complete candidates can be evaluated reliably. Many vision-language tasks instead provide only partial verification: a finding, span, va...

📖 Read original article


71. The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess Reasoning ​

Author: Ishan S. Kshirsagar
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.20952v1 Announce Type: new Abstract: Latent, or silent, reasoning lets language models carry out intermediate computation in continuous vector space instead of words, and is widely assumed to function as an internal scratchpad the model actively consults during inference. Whether that ass...

📖 Read original article


72. ADABORD: a novel AdaBoost approach for ordinal classification ​

Author: Rafael Ayll'on-Gavil'an, Francisco Jos'e Mart'inez-Estudillo, David Guijo-Rubio, C'esar Herv'as-Mart'inez, Pedro A. Guti'errez
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21003v1 Announce Type: new Abstract: Ordinal Classification (OC) deals with classification tasks where the classes follow a natural order. Despite the progress in OC, many existing approaches fail to fully leverage the ordinal information, treating the problem as nominal classification an...

📖 Read original article


73. Weight-norm Criticality: A Mechanism for Loss Spikes Induced by the Normalization and Weight Decay ​

Author: Xiaolong Li, Zhangchen Zhou, Zhi-Qin John Xu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2607.21005v1 Announce Type: new Abstract: Most explanations of training instability focus on \emph{learning-rate criticality}, typically characterized by the Edge of Stability, beyond which optimization becomes unstable. We argue that, in practical deep neural network training, there is an add...

📖 Read original article


74. Regularized Optimization on Grassmann Manifold: Theory, Algorithm and Applications ​

Author: Zhuan Liang, Zheng Zhai
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21039v1 Announce Type: new Abstract: Spectral methods are among the most widely used techniques for community detection, clustering, and graph learning. Their performance, however, critically depends on the accurate estimation of the underlying spectral subspace and can deteriorate substa...

📖 Read original article


75. Counterfactual Explainability Framework With CycleGAN And Counterfactual-Classifier Alignnment Score for Retinal Disease Classification ​

Author: Kritanu Chattopadhyay, Sayanjit Singha Roy, Soumya Chatterjee
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.21068v1 Announce Type: new Abstract: Automated detection of vision impairing retina-based ocular conditions from fundus images is important for early screening, timely referral and reducing dependency on specialist-only assessment, for which neural network-based deep learning (DL) models ...

📖 Read original article


76. From Evaluation to Optimisation: Hierarchy-Aware Training Signals for CWE Prediction in Python ​

Author: Muntasir Adnan, Manile Srun, Carlos C. N. Kuhn
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21069v1 Announce Type: new Abstract: The original ALPHA benchmark introduced a taxonomy-aware penalty for evaluating CWE-level vulnerability prediction in Python and proposed that the penalty could theoretically also serve as a training signal. This paper provides that validation. We comp...

📖 Read original article


77. Spectral Transformation for Layer-wise Global Rank Discovery in Federated LoRA for Vision Transformers ​

Author: Hariharan Ramesh, Jyotikrishna Dass
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2607.21074v1 Announce Type: new Abstract: Fine-tuning Vision Transformers (ViTs) with low-rank adapters (LoRA) promises better communication efficiency under federated setup, yet existing aggregation strategies face fundamental limitations. Independently averaging these LoRA factors is mathema...

📖 Read original article


78. Nipping the Butterfly Effect in the Bud: Self-Output Fine-Tuning for Autoregressive Weather Prediction ​

Author: Yun-Ye Cai, Hsuan-Tien Lin
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21080v1 Announce Type: new Abstract: Long-horizon weather forecasting is a fundamental challenge in atmospheric science, for which autoregressive Deep Learning Weather Prediction (DLWP) has emerged as the primary paradigm. Although the autoregressive pipeline is highly scalable and flexib...

📖 Read original article


79. CASC: Causal Adversarial Subspace Clustering for Multivariate Spatiotemporal Data ​

Author: Francis Ndikum Nji, Vandana Janeja, Jianwu Wang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21088v1 Announce Type: new Abstract: Deep subspace clustering plays a critical role in applications involving multivariate spatiotemporal data, such as sea ice monitoring, disease spread analysis, and tracking neuro-degeneration over time. Despite recent advances, existing methods primari...

📖 Read original article


80. Training Large Language Models for Self-Explanation Faithfulness ​

Author: Yeoktatt Cheah, Mar'ia P'erez-Ortiz, Noah Y. Siegel, Oana-Maria Camburu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.21090v1 Announce Type: new Abstract: We propose a Reinforcement Learning (RL) method to directly optimize the faithfulness of self-explanations - the extent to which a model's generated reasoning accurately reflects its internal decision-making process. While existing work focuses on eval...

📖 Read original article


81. A Polynomial Architecture-Attribution Co-Design Framework for Exact Aumann-Shapley Attribution in GNNs ​

Author: Bizu Feng, Zhimu Yang, Shuming Wang, Shaode Yu, Yuan Cheng, Xiaojun Qian, Zixin Hu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21094v1 Announce Type: new Abstract: We study feature-level and node-level explanations for graph neural networks (GNNs) through the lens of Aumann-Shapley attribution. Path-integral methods such as Integrated Gradients provide an axiomatic formulation of attribution, but their practical ...

📖 Read original article


82. Smooth Neural Point Processes via B-Splines ​

Author: Michele Bellomo, Riccardo Ramaschi, Alberto Dolara, Tomaso Aste
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.21098v1 Announce Type: new Abstract: Temporal point processes (TPPs) provide a general and flexible framework for modeling sequences of events in continuous time. Neural networks have been successfully employed to model TPPs in a highly expressive and data-driven way. Neural TPPs are typi...

📖 Read original article


83. TOUR: A Trajectory-Level Unlearning Benchmark for Offline Reinforcement Learning ​

Author: Chaofan Pan, Lingfei Ren, Xiangyu Jiang, Yanhua Li, Xuemei Cao, Xiangkun Wang, Hao Yu, Wei Wei, Xin Yang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.21111v1 Announce Type: new Abstract: Offline Reinforcement Learning (RL) agents are trained on fixed behavioral trajectories, which makes trajectory-level deletion important when selected data must be removed after training. Evaluating such deletion is difficult because a lower membership...

📖 Read original article


84. GlucoTune: A Unified Framework for Blood Glucose Preprocessing, Forecasting, and Benchmarking in Diabetes ​

Author: Davide Marelli, Giorgia Rigamonti, Mirko Paolo Barbato, Paolo Napoletano
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.21117v1 Announce Type: new Abstract: Preprocessing blood glucose time-series data is a critical yet often overlooked step in developing data-driven methods for diabetes management, particularly for type 1 diabetes. The lack of standardized preprocessing workflows and evaluation protocols ...

📖 Read original article


85. Relative Value Learning ​

Author: Marc H"oftmann, Jan Robine, Stefan Harmeling
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.21120v1 Announce Type: new Abstract: In reinforcement learning, critics typically estimate absolute state values $V(s)$, estimating how good a particular situation is in isolation. However, it turns out that only differences in value are relevant for control. Motivated by this, we propose...

📖 Read original article


86. Demographically-Informed Heat-Mortality Risk Curves via Risk Graph Neural Networks ​

Author: Alex O. Davies, Eunice Lo, Rui Zhu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.21131v1 Announce Type: new Abstract: Estimating heat-related mortality risk is a core task in environmental epidemiology, typically addressed with Distributed Lag Non-linear Models (DLNMs); interpretable exposure-response surfaces fitted to temperature-mortality time series. DLNMs are eff...

📖 Read original article


87. Agree on the Model, Verify the Inference: GKR Protocols for HND-Based Transformer Inference ​

Author: Xiaolong Liang, Juanjuan Li, Rui Qin, Yisheng Lv
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2607.21162v1 Announce Type: new Abstract: Outsourced Transformer inference exposes clients to model substitution and incomplete execution, while direct replay removes the computational benefit of delegation. We present GKR-HND, a registered-model protocol for verifying the polynomial backbone ...

📖 Read original article


88. Automated Synthesis and Adversarial Validation of Executable Causal Research Pipelines ​

Author: Irena Girshovitz, Dan Zeltzer, Ran Gilad-Bachrach
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21173v1 Announce Type: new Abstract: While automated research systems promise to accelerate empirical analysis, they are prone to silent failures: instances in which analysis code executes successfully yet relies on invalid causal assumptions. We present the Artificial Intelligence (AI)-b...

📖 Read original article


89. Filter Learning for Subgraphs: Algebras and Performance Risk Bounds ​

Author: Purui Zhang, Feng Ji, Yanan Zhao, Bihan Wen, Wee Peng Tay
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2607.21263v1 Announce Type: new Abstract: Graph signal processing tasks that leverage spectral information typically assume access to the complete graph topology, which is often unavailable in practice. We propose a systematic framework for subgraph filter learning (SFL), where subgraph-suppor...

📖 Read original article


90. The Dark Room in the Reward Channel: Dense Prediction Rewards Collapse GRPO-Trained LLM Agents -- and What Actually Works ​

Author: Yu Wang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21273v1 Announce Type: new Abstract: Dense per-step supervision is an appealing remedy for sparse-reward, long-horizon LLM agents: reward the agent for predicting its next observation, and memory should follow. We show that under group-normalized RL (GRPO), this recipe does not merely fai...

📖 Read original article


91. Multi-Task Learning for Heterogeneous Prediction from Video Game State with Transfer Learning ​

Author: Jonas Pech'e, Aliaksei Tsishurou, Alexander Zap, G"unter Wallner
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.21290v1 Announce Type: new Abstract: Multi-task learning (MTL) is a promising approach for prediction tasks derived from video game state data, as modern game telemetry provides multiple related supervision signals from the same structured observations. We study whether a shared model tra...

📖 Read original article


92. AI Assistants Overassist ​

Author: Verona Teo, Raghav Jain, Tobias Gerstenberg, Max Kleiman-Weiner
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL, cs.CY, cs.HC

arXiv:2607.21306v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as tutors and thought partners, helping users reason through problems. While guidance from AI assistants can scaffold thinking and foster learning, such benefits depend on how they help--for instance, ...

📖 Read original article


93. M$^3$-Gen: Interpretable Multimodal Generation of Gene Expression Profiles Using Clinical and Imaging Data ​

Author: Francesca Pia Panaccione, Carlo Sgaravatti, Marco Venere
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.21343v1 Announce Type: new Abstract: Integrating heterogeneous biomedical data, including clinical metadata, histopathology images, and molecular profiles, is crucial for comprehensive disease understanding. However, gene expression data acquisition remains constrained by high costs and p...

📖 Read original article


94. How Many Bits Can an Adapter Write? Measuring the Capacity and Memorization of Parameter-Efficient Fine-Tuning ​

Author: Kaizhen Tan, Heqing Du, Yang Feng
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21351v1 Announce Type: new Abstract: A LoRA adapter is a few megabytes that almost everyone treats as a skill rather than a record of the data behind it. We put that assumption on a scale. Extending compression-based memorization analysis to the frozen-base setting, we measure directly, i...

📖 Read original article


95. Gradient Concentration, Not Weight Saliency, Explains Representation-Level Class Unlearning ​

Author: Billel Habbati, Alessio Merlo, Luca Verderame, Meriem Guerar
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21353v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific training data while preserving model utility. Many state-of-the-art approaches pursue this goal by restricting the forgetting update to a subset of parameters selected through gradient-based s...

📖 Read original article


96. Emergent Misalignment Recruits a Pre-existing Persona Subspace ​

Author: Mohammed Suhail B Nadaf
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21356v1 Announce Type: new Abstract: Fine-tuning an aligned language model on a narrow stream of bad advice can make it broadly misaligned on questions unrelated to the training data, a phenomenon called emergent misalignment. We ask why the narrow lesson generalizes at all, and we find t...

📖 Read original article


97. Hilbert Operator for Progressive Encoding (HOPE): A Mathematical Framework for Deconstructing Learned Representations in Deep Networks ​

Author: Hossein Mobahi, Peter L. Bartlett
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.21366v1 Announce Type: new Abstract: Deep neural networks encode complex representations, but deconstructing this internal knowledge remains a challenge. Given the link between learning and compression, network compression offers a promising lens to analyze this knowledge. However, standa...

📖 Read original article


98. Mean-to-Score Discrete Diffusion: Posterior-Mean Denoisers for Score Entropy ​

Author: Jingyuan Li, Xiaoyi Jiang, Yixuan Jiang, Wei Liu, Yi Zhu, Zuoqiang Shi, Pipi Hu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.21372v1 Announce Type: new Abstract: Score Entropy Discrete Diffusion (SEDD) parameterizes discrete reverse processes with unconstrained positive score ratios. While positivity guarantees nonnegative reverse jump rates, it does not ensure Bayes realizability: ratios at a noisy state need ...

📖 Read original article


99. A Diffusion-Model Subpopulation Digital Twin for Mobile Health Deployment: A Case Study on the HeartSteps Intervention ​

Author: Ziping Xu, Yuyi Chang, Chenshun Ni, Nithin Sugavanam, Asim H. Gazi, Pedja Klasnja, Emre Ertin, Susan A. Murphy
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ME

arXiv:2607.21403v1 Announce Type: new Abstract: Mobile-health interventions increasingly use online learning and decision making algorithms to personalize when to nudge users toward healthier behavior, but a poorly designed algorithm can burden and disengage participants. New algorithm design decisi...

📖 Read original article


100. Semantic-Aware Task Clustering for Constructive and Cooperative Multi-Tasking ​

Author: Ahmad Halimi Razlighi, Maximilian H. V. Tillmann, Edgar Beck, Bho Matthiesen, Armin Dekorsy
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, eess.SP, math.IT, stat.ML

arXiv:2607.21426v1 Announce Type: new Abstract: Cooperative multi-task semantic communication (CMT-SemCom) improves task execution performance by leveraging shared representations. However, as we demonstrated in [1], cooperative multi-tasking can be either constructive or destructive, depending on t...

📖 Read original article


101. Context-weighted Discrete Flow Matching ​

Author: Daniil Cherniavskii, Daniel Severo, Karen Ullrich
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21427v1 Announce Type: new Abstract: Discrete flow matching provides a flexible framework for generative modeling on discrete structures. However, the standard factorized training objective exposes the model to targets of varying difficulty, mixing well-conditioned, predictable tokens wit...

📖 Read original article


102. KroQuant: Kronecker-Structured Block Transforms for Efficient Post-Training Quantization of Diffusion Transformers ​

Author: Yann Bouquet, Alireza Khodamoradi, Kristof Denolf, Mathieu Salzmann
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.21446v1 Announce Type: new Abstract: Post-training quantization (PTQ) of diffusion transformers (DiTs) to W4A4 severely degrades output quality, because activations entering each linear layer contain outliers that 4-bit formats cannot represent. The standard fix applies an invertible line...

📖 Read original article


103. Test-Time Scaling via Error Localization ​

Author: Rajiv Shailesh Chitale, Rahul Madhavan, Taneesh Gupta, Deepanway Ghosal, Aravindan Raghuveer
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21453v1 Announce Type: new Abstract: Scaling inference-time computation has emerged as a reliable method to improve the performance of large language models on complex reasoning and programming tasks. However, standard approaches such as independent sampling and sequential multi-turn refi...

📖 Read original article


104. Error Certificates for KV-Cache Eviction via Randomized Design ​

Author: Peng Xie
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.21475v1 Announce Type: new Abstract: Deterministic KV-cache eviction keeps the top-$k$ tokens under an importance score and deletes the rest. We prove that this design cannot know what it destroyed: evicted values can be altered so that everything the serving system retains is unchanged w...

📖 Read original article


105. Finite-Sample Coverage Audits for High-Recall Candidate Generation: Certification and Learning-Theoretic Design ​

Author: Martin Anthony, Kaveh Salehzadeh Nobari
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.21480v1 Announce Type: new Abstract: An initial high-recall stage in an empirical pipeline decides which items pass to later review, labelling, or modelling, and relevant items it misses are lost to every subsequent stage. We study how many audit labels are needed to certify, with finite-...

📖 Read original article


106. Compact Latent Coordination for Autonomous Vehicles at Unsignalized Intersections ​

Author: Gil Lifshits, Igal Bilik, Gilad Katz
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA, cs.RO

arXiv:2607.21488v1 Announce Type: new Abstract: Coordinating autonomous vehicles at unsignalized intersections remains a critical challenge for multi-agent reinforcement learning (MARL) systems, which typically struggle with combinatorial action spaces, reliance on privileged information, or rigid a...

📖 Read original article


107. Windowed-MTP: Removing the Full-Context Draft-KV Tax at Million-Token Context ​

Author: Alagappan Valliappan
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.PF

arXiv:2607.21535v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by having a cheap draft propose tokens that a target verifies in parallel. Frontier models increasingly ship a built-in Multi-Token-Prediction (MTP/NEXTN) draft head under the assumption that t...

📖 Read original article


108. Zero-Flow Two-Sample Tests ​

Author: Yakun Wang, Leyang Wang, Song Liu, Taiji Suzuki
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.21542v1 Announce Type: new Abstract: We propose a new approach to two-sample testing for deciding whether two sets of samples are drawn from the same distribution. The test is built on a statistical discrepancy based on the zero-flow criterion, termed zero-flow discrepancy (ZFD). We prove...

📖 Read original article


109. X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment ​

Author: Dongjie Fu, Di Cao, Xize Cheng, Zihan Zhang, Wenxu Jia, Yifu Chen, Shengpeng Ji, Yu Zhang, Tao Jin
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21550v1 Announce Type: new Abstract: While large audio-language models have achieved remarkable progress in auditory perception, they still lag behind text-based large language models in deep logical reasoning, primarily due to the scarcity of high-quality audio reasoning data. To bridge ...

📖 Read original article


110. Graph Learning on Ensembles of Cyclic Peptides: An Investigation of Molecular Ensemble Modeling ​

Author: Aaron Feller, Kris Deibler, Maxim Secor
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM

arXiv:2607.21561v1 Announce Type: new Abstract: Molecular property prediction from structure often uses a single representative conformation, even though many molecules exist as conformational ensembles in solution. We introduce EnsembleEGNN, a molecular ensemble foundation model that encodes an ens...

📖 Read original article


111. Beyond Sufficiency: Time Series Explanation with Counterfactual Necessity ​

Author: Hongnan Ma, Yiwei Shi, Mengyue Yang, Weiru Liu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.21573v1 Announce Type: new Abstract: Faithful explanations of time-series classifiers should identify subsequences that are not only sufficient to preserve a black-box model's prediction, but also necessary for maintaining it. However, existing sufficiency-oriented methods can assign high...

📖 Read original article


112. Expanding Flow Maps ​

Author: Sophia Tang, Pranam Chatterjee
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.21585v1 Announce Type: new Abstract: Flow-based generative models have enabled remarkable progress in fast and controllable generation across continuous and discrete state spaces, yet existing parameterizations are constrained to fixed dimensions or fixed sequence lengths. Here, we introd...

📖 Read original article


113. Deblurring in the Wild: A Real-World Image Deblurring Dataset from Smartphone High-Speed Videos ​

Author: Syed Mumtahin Mahmud, Mahdi Mohd Hossain Noki, Prothito Shovon Majumder, Abdul Mohaimen Al Radi, Sudipto Das Sukanto, Afia Lubaina, Md. Mosaddek Khan
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2506.19445v4 Announce Type: cross Abstract: We introduce the largest real-world image deblurring dataset constructed from smartphone slow-motion videos. Using 240 frames captured over one second, we simulate realistic long-exposure blur by averaging frames to produce blurry images, while using...

📖 Read original article


114. From Attention to Frequency: Integration of Vision Transformer and FFT-ReLU for Enhanced Image Deblurring ​

Author: Syed Mumtahin Mahmud, Mahdi Mohd Hossain Noki, Prothito Shovon Majumder, Abdul Mohaimen Al Radi, Md. Haider Ali, Md. Mosaddek Khan
Published: 7/24/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.CV, cs.LG

arXiv:2511.10806v1 Announce Type: cross Abstract: Image deblurring is vital in computer vision, aiming to recover sharp images from blurry ones caused by motion or camera shake. While deep learning approaches such as CNNs and Vision Transformers (ViTs) have advanced this field, they often struggle w...

📖 Read original article


115. Preference Tuning as Spectral Update Reorganization ​

Author: Peiyan Zhang, Haibo Jin, Liying Kang, Haohan Wang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.20438v1 Announce Type: cross Abstract: Preference-based post-training is usually understood through endpoint behavior, yet the learned update that produces this behavior remains largely opaque. We study RLHF and related preference optimization through the spectral structure of their induc...

📖 Read original article


116. Belief Propagation in LLM World Models: Measuring Strategic Information Bias with Prediction Markets ​

Author: Mykola Khandoga, Yevhen Kostiuk, Anton Polishko, Yurii Filipchuk, Kostiantyn Kozlov, Dmytro Zamriy, Artur Kiulian
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.20441v1 Announce Type: cross Abstract: Every information ecosystem produces beliefs that shape strategic decisions. Both human analysts and AI systems inherit the blind spots of their information sources. We show that LLMs, combined with prediction markets, function as a calibrated instru...

📖 Read original article


117. Naver-News-KO: A Korean News Summarization Dataset for Open-Source Fine-Tuning of Summarization Models ​

Author: Daekeun Kim
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.20442v1 Announce Type: cross Abstract: We release Naver-News-KO, a Korean news summarization dataset of 27,400 (document, summary) pairs collected from Naver News over a ten-day window in July 2022 across two categories (Economy and IT/Science; 77/23 split), with train/validation/test par...

📖 Read original article


118. GLAN-QnA-KR: A Seedless Taxonomy-Driven Korean Instruction Corpus ​

Author: Daekeun Kim
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.20443v1 Announce Type: cross Abstract: We release GLAN-QnA-KR, a 303,581-row openly redistributable Korean instruction-QA corpus produced via the seedless taxonomy-driven GLAN synthesis pipeline with Microsoft's Phi-3.5-MoE-instruct as the producer model (generation: 2024-12; release: 202...

📖 Read original article


119. thaulab@EEUCA 2026: Who Said What to Whom? A Targeting-Aware Neural-Symbolic Pipeline for Gaming Toxicity Detection ​

Author: Anmol Guragain, Marcos Estecha-Garitagoitia, Luis Fernando D'Haro Enr'iquez, Ricardo de C'ordoba
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.20447v1 Announce Type: cross Abstract: This paper describes our system for the EEUCA 2026 Shared Task on toxicity classification in gaming chat. We implement a three-stage pipeline combining an ensemble of two compact transformers (DeBERTa-v3-base, 184M; XLM-RoBERTa-base, 278M) with a Lin...

📖 Read original article


120. Domyn-Small: A European 10B Reasoning Language Model ​

Author: Simone Angarano, Francesco Bertolotti, Federico D'Ambrosio, Michele Resta, Alessandro Rognoni, Nicol`o Ruggeri, Dario Salvati, Andrea Valenti, Alberto Veneri, Martin Cimmino
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.20448v1 Announce Type: cross Abstract: We introduce Domyn-Small, a 10-billion-parameter open-weight reasoning language model released under the MIT license. Domyn-Small is the product of an initial pre-training phase on 9 trillion tokens multilingual data, followed by a post-training pipe...

📖 Read original article


121. A Knowledge-Injection Framework for Zero-Shot Adaptation of LLMs to Delirium Prediction ​

Author: Jessica Sena, Shesadree Priyadarshani, Miguel Contreras, Bharat Gandhi, Scott Siegel, Subhash Nerella, Parisa Rashidi
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.20453v1 Announce Type: cross Abstract: Large language models show promise for clinical prediction, but zero-shot performance on specialized tasks is limited by incomplete domain knowledge, especially for smaller locally deployable models. We present a lightweight knowledge-injection frame...

📖 Read original article


122. DC-Leap: Training-Free Acceleration of dLLMs via Draft-Guided Contiguous Leaping Decoding ​

Author: Yanhua Jiao, Tianyi Wu, Xiaoxi Sun, Yulin Li, HuiLing Zhen, Libo Qin, Baotian Hu, Zhuotao Tian, Min Zhang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.20467v1 Announce Type: cross Abstract: While parallel decoding is central to the efficiency of Diffusion Large Language Models (dLLMs), current strategies are often hindered by overly conservative confidence thresholds. These thresholds, necessitated by the Joint Probability Dependence Er...

📖 Read original article


123. SonicSampler: Unified Tile-Aware Kernels for LLM Sampling and Speculative Verification ​

Author: Pragaash Ponnusamy, Shivam Sahni, Jue Wang, Tri Dao
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.20475v1 Announce Type: cross Abstract: Sampling in LLM inference comprises a combinatorial set of logit processing, token selection, and verification operations for speculative decoding. However, existing implementations either accelerate only subsets of this pipeline, rely on multiple ke...

📖 Read original article


124. Expectation Alignment of Language Models for Real-World User Expectations ​

Author: Miaomiao Li, Yang Wang, Bin Liang, Shudong Liu, Zhiwei Zhang, Kam-Fai Wong
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.20485v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated remarkable performance on standard benchmarks, yet it remains largely unexplored whether they truly meet user expectations. Existing evaluation approaches, relying on model heuristics, expert rubrics, or...

📖 Read original article


125. DFAH-Bench: Benchmarking Observable Agent Instability in Financial Decision-Making ​

Author: Raffi Khatchadourian
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.20491v1 Announce Type: cross Abstract: Standard evaluation benchmarks measure what a tool-using agent decides, not whether it arrives at that decision through the same process each time. We introduce DFAH-Bench, a replay benchmark that measures observable behavioral instability in financi...

📖 Read original article


126. Attention-based Experience Replay Framework for Continual Learning of Agnostic Time Series Forecasting Models ​

Author: Quentin Besnard (RFAI), Nicolas Ragot (RFAI)
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.20493v1 Announce Type: cross Abstract: Deep learning has led to remarkable progress in artificial intelligence, particularly in robotics, imaging and sound processing. However, a major limitation of neural networks remains their strong dependence on large and stationary datasets. In many ...

📖 Read original article


127. Isolating LLM Alignment from Regex: Zero Coverage and Metric-Dependent Divergence Under Adversarial Mutation ​

Author: Alexandre Cristov~ao Maiorano
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CR, cs.LG

arXiv:2607.20494v1 Announce Type: cross Abstract: Production LLM applications commonly stack a regex filter in front of model-side alignment; prior work found no measurable coverage gain from adding a live Gemini backend behind an active regex filter. We ask whether that ceiling holds when the corpu...

📖 Read original article


128. LeanFlow: A Case Study in Workflow-Driven Lean Autoformalization ​

Author: Lazar Milikic, Simon Guilloud, Khanh Nguyen, Viktor Kuncak
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO

arXiv:2607.20503v1 Announce Type: cross Abstract: We present and evaluate LeanFlow, an LLM agent system specialized for translating mathematical papers into buildable Lean projects. Recent verifier-in-the-loop systems show that large formal artifacts can be produced, but it remains unclear which run...

📖 Read original article


129. HERMES: Heterogeneous Edge-Relational Multi-Head Embedded SSM Attention for Traffic Conflict Prediction at Signalized Intersections ​

Author: Md Monzurul Islam, Subasish Das
Published: 7/24/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.20505v1 Announce Type: cross Abstract: Surrogate safety measures (SSMs) enable proactive traffic safety assessment, but many existing methods evaluate pairwise interactions independently or flatten multi-agent scenes into fixed feature vectors, limiting their ability to represent heteroge...

📖 Read original article


130. MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference ​

Author: Jingquan Chen, Jinghua Piao, Jie Feng, Shaogang Hu, Yong Li
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.20507v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for program-aided reasoning, agentic decision making, and structured task execution, but these applications often incur high inference cost. We present MiniCache, a reusable program caching framework...

📖 Read original article


131. ConfidenceBench: Evaluating Confidence Calibration in Large Language Models ​

Author: Matthew ffrench-Constant, Daniel Yang, Xinmeng Huang, Sanyam Kapoor
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, stat.ML

arXiv:2607.20526v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in settings where fluent but incorrect answers can be costly. In these settings, accuracy alone is insufficient: models must also know when they are likely to be wrong. We present ConfidenceBench...

📖 Read original article


132. AppWorld-UL: Benchmarking Diverse Agent-User Interactions for Tool-Use ​

Author: Junzhi Chen, Harsh Trivedi, Jane Pan, Michael JQ Zhang, Tejas Srinivasan, Niranjan Balasubramanian, Ashish Sabharwal
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2607.20536v1 Announce Type: cross Abstract: Tool-use agents that address day-to-day digital tasks such as ordering groceries must not only operate applications, but also interact with the user, e.g., to ask clarification questions, prompt for confirmation, and inform the user when the instruct...

📖 Read original article


133. AI-Driven Multi-Hop Relay Selection for Smart Urban NR-V2X Networks via Learning-to-Optimize Graph Neural Networks ​

Author: Giambattista Amati, Federica Mangiatordi, Simone Angelini, Emiliano Pallotti, Pierpaolo Salvo
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.NI

arXiv:2607.20554v1 Announce Type: cross Abstract: Reliable and low-latency NR-V2X communications are essential for smart mobility in dense urban environments. However, limited Road-Side Unit (RSU) density, frequent non-line-of-sight conditions, and highly dynamic vehicular topologies often prevent m...

📖 Read original article


134. Machine-Learned Compact Subspace Generation for Quantum Selected Configuration Interaction within Density Matrix Embedding Framework ​

Author: Ashish Kumar Patra, Anurag K. S. V., Ruchika Bhat, Sai Shankar P., Rahul Maitra, Jaiganesh G
Published: 7/24/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG, physics.chem-ph, q-bio.BM

arXiv:2607.20585v1 Announce Type: cross Abstract: Sample-based Quantum Diagonalization (SQD), an extension of Quantum Selected Configuration Interaction (QSCI), has emerged as a promising hybrid quantum-classical paradigm for computing molecular ground state energies. By leveraging quantum sampling ...

📖 Read original article


135. SPECTRA: State-Space Exogenous Context and Temporal-Frequency Resolution Architecture for Probabilistic Energy Forecasting ​

Author: Hang Ye, Xinyan Jiang, Yuedong Shi, Yangxin Zhu, Jianming Wei, Tian Zheng, Xiaoying Zheng, Yongxin Zhu
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.20587v1 Announce Type: cross Abstract: Modern power systems increasingly require probabilistic forecasts amid interacting uncertainties from renewable intermittency, flexible demand, market volatility, and weather-dependent generation. However, existing methods often treat multi-scale dec...

📖 Read original article


136. Algorithmic Approaches to Sequential Decision-Making and Social Epistemology ​

Author: Kavya Ravichandran
Published: 7/24/2026, 4:00:00 AM
Categories: cs.DS, cs.LG

arXiv:2607.20636v1 Announce Type: cross Abstract: As humans, we face many decisions that require us to choose between sticking to something and giving up. This thesis uses algorithmic tools to derive insights about such decision-making problems in theoretical models, studying both near-optimal metho...

📖 Read original article


137. Masked Topology Modeling for Self-Supervised Learning on Parametric CAD ​

Author: Heinrich Jiang, Jennifer Jang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.20642v1 Announce Type: cross Abstract: Computer aided design (CAD) is ubiquitous: virtually any modern object was designed using editable CAD tools. However, with the shortage of available CAD datasets in its native editable and parametric format, boundary representation (B-Rep), it is ev...

📖 Read original article


138. Frontier Financial Judgement: Can agents tell what might move a stock? ​

Author: Joshua Harris
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.20645v1 Announce Type: cross Abstract: We introduce Frontier Financial Judgement, a challenging new benchmark developed in collaboration with professional equity analysts to assess agents' ability to replicate expert human judgements. Rapidly identifying new information, evaluating its im...

📖 Read original article


139. PhysCoRe: Physics-Corrected Residual World Models for Material-Aware Deformable Dynamics ​

Author: Haocheng Yin, Shuohan Tao, Yongsheng Chen, Lu Gan
Published: 7/24/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2607.20653v1 Announce Type: cross Abstract: Predicting how deformable objects evolve under robotic manipulation is a longstanding challenge. Existing approaches typically rely on per-object optimization to fit material parameters, which can be slow and cannot generalize, while end-to-end learn...

📖 Read original article


Author: Jack Beda, Djordje Mihajlovic, Kasturi Barkataki, Davide Michieletto
Published: 7/24/2026, 4:00:00 AM
Categories: math.GT, cond-mat.soft, cs.LG

arXiv:2607.20657v1 Announce Type: cross Abstract: Unique and rapid classification of knots and links is an open mathematical problem that is relevant to a range of (bio)physical systems, including polymer melts, DNA, and proteins. In this paper, we explore a data-driven approach to the classificatio...

📖 Read original article


141. Cross-Domain Generalization in Optical Networks via Joint Contrastive and Classification Learning ​

Author: Ali Al Housseini, Carlos Natalino, Paolo Monti, Omran Ayoub
Published: 7/24/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2607.20666v1 Announce Type: cross Abstract: The robustness of machine learning techniques across heterogeneous network domains remains an open challenge in optical networks. Models trained on data from a specific topology or operational configuration often exhibit degraded performance when dep...

📖 Read original article


142. Leaky Language Models: Stealing Architecture and Inference Optimizations via Per-Token Timing ​

Author: Sadegh Majidi, Niloofar Mireshghallah, Kazem Taram
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2607.20723v1 Announce Type: cross Abstract: This work presents LeakyLMs, a set of attacks that leak proprietary model, architecture, and deployment information from production language models. LeakyLMs is the first to demonstrate that key model and deployment details can be inferred using only...

📖 Read original article


143. Pipelined Gradient Coding ​

Author: Xian Su, Jun Li
Published: 7/24/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT

arXiv:2607.20739v1 Announce Type: cross Abstract: In large-scale machine learning, distributed training commonly involves multiple workers evaluating the gradients of the model on different dataset partitions. A common challenge is the presence of straggling workers, which may significantly slow dow...

📖 Read original article


144. Self-Supervised Bio-Inspired Robotic Trajectory Planning with Obstacle Avoidance ​

Author: Miroslav Krupa, Miroslav Cibula, Krist'ina Malinovsk'a
Published: 7/24/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.20743v1 Announce Type: cross Abstract: Trajectory planning is a fundamental problem in robotics, requiring the generation of collision-free and efficient trajectories in a potentially complex environment. While sampling-based planners remain the dominant approach, they are often computati...

📖 Read original article


145. Are Diversity Metrics Measuring Diversity? A Capability-Controlled Audit of Majority-Vote Gain in LLM Ensembles ​

Author: Donghwan Kim
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.20768v1 Announce Type: cross Abstract: Majority voting over LLMs is widely assumed to benefit from diversity, and diversity measures are used to choose which models to combine. We ask whether five such measures track diversity or mainly re-express capability, auditing them as predictors o...

📖 Read original article


146. Emergent Compositional Skills in Mixture-of-Experts VLAs ​

Author: Shlok Shah, Rhiaan Jhaveri, Tharun Kumar Tiruppali Kalidoss, Chirayu Nimonkar, Ishaan Javali
Published: 7/24/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.20771v1 Announce Type: cross Abstract: We consider the problem of learning compositional robot policies end-to-end from expert demonstrations, without any pre-specified notion of task decomposition or hierarchy. We ask whether a VLA trained with a simplified Mixture-of-Experts (MoE) actio...

📖 Read original article


147. The Geometry of Personality: Activation Steering with Jungian Cognitive Functions ​

Author: Liu Zai (University of Glasgow), Yumeng Wang (Leiden University), Junchen Fu (University of Glasgow), Joemon M. Jose (University of Glasgow)
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.20803v1 Announce Type: cross Abstract: Activation steering enables control and interpretation of LLMs, yet existing work primarily models personality through static trait frameworks such as the Big Five. We investigate whether personality can instead be represented and controlled as a set...

📖 Read original article


148. Probabilistic Residual Learning for Online Recommendations ​

Author: Wenyuan Wang, Yusong Zhao, Zihao Xu, Hengyi Wang, Qi Xu, Zhigang Hua, Yan Xie, Yi Wang, Zihao Zhao, Bo Long, Chengzhi Mao, Shuang Yang, Hengguan Huang, Hao Wang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.IR, cs.AI, cs.LG

arXiv:2607.20863v1 Announce Type: cross Abstract: Modern recommender systems are typically based on deep learning (DL) models, where a dense encoder learns representations of users and items. As a result, these systems often suffer from the black-box nature and computational complexity of the underl...

📖 Read original article


149. Machine Learning for Charge State Characterization of Isolated Double Quantum Dots ​

Author: Hyma Vallabhapurapu, Marco Candido, Krishna Choudhary, Paul Steinacker, Ensar Vahapoglu, Chris Escott, Wee Han Lim, Andre Saraiva, Nard Dumoulin Stuyck, MengKe Feng
Published: 7/24/2026, 4:00:00 AM
Categories: cond-mat.mes-hall, cs.LG

arXiv:2607.20871v1 Announce Type: cross Abstract: Scaling semiconductor quantum dot arrays toward fault-tolerant quantum computing requires efficient tuneup of spin qubits, a process that depends on the analysis of charge stability maps (CSMs) and remains largely manual. While machine learning has b...

📖 Read original article


150. RadioTrace: Transmitter-Aware Diffusion for Radio Map Estimation without Deployment-Time Fine-Tuning ​

Author: Liu Yang, Qiang Li, Zhuo Cao, Weijie Xiong, Guomin Sun, Jingran Lin
Published: 7/24/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2607.20909v1 Announce Type: cross Abstract: Radio map (RM) estimation aims to reconstruct the spatial distribution of wireless signal characteristics, such as received signal strength (RSS), from sparse measurements, a task that is critical for spectrum management, interference mitigation, and...

📖 Read original article


151. An Analytically Trained Variational Surrogate for Quantum Phase Estimation on NISQ Hardware ​

Author: Mousumi Kundu, Ashish Kumar Patra, Anurag K. S. V., Ruchika Bhat, Sai Shankar P., Alok Shukla, Jaiganesh G
Published: 7/24/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG, physics.chem-ph

arXiv:2607.20943v1 Announce Type: cross Abstract: Quantum Phase Estimation (QPE) is a foundational algorithm for molecular ground-state energy estimation, but its deep circuit requirements make direct hardware execution impractical on Noisy Intermediate-Scale Quantum (NISQ) devices. We present an an...

📖 Read original article


152. Automatic knot selection in smooth additive models ​

Author: Nicol'as Carrizosa, Vanesa Guerrero, Mar'ia Durb'an
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.21083v1 Announce Type: cross Abstract: B-spline regression constitutes a widely used framework for nonparametric modeling. The performance of this methodology depends on specifying the number and placement of changepoints, known as knots, prior to the estimation process. Such knot sequenc...

📖 Read original article


153. Approximate Quantum State Preparation Through Proximal Policy Optimization ​

Author: Marco Mordacci, Michele Amoretti
Published: 7/24/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG

arXiv:2607.21121v1 Announce Type: cross Abstract: In this work, a quantum architecture search framework for approximate quantum state preparation (QSP) is proposed. QSP is a challenging task, since the search space grows exponentially with the number of qubits, making the identification of the optim...

📖 Read original article


154. Safety-oriented sidewalk and road segmentation for smartphone-based assistive navigation ​

Author: Hakan Calim, Anamaria Dumitrescu, Adarsh Bhandary Panambur, Huzaifa Asif, Andreas Maier
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.21137v1 Announce Type: cross Abstract: Independent sidewalk mobility is essential for blind and visually impaired pedestrians (BVIPs), yet smartphone-based assistive navigation requires perception models that distinguish walkable sidewalks from adjacent unsafe regions. This study presents...

📖 Read original article


155. Explaining Weather Bulletins via ILP ​

Author: Enrico Santi (University of Udine, DMIF), Alessandro Dal Pal`u (University of Parma, SMFI), Agostino Dovier (University of Udine, DMIF), Talissa Dreossi (University of Udine, DMIF), Andrea Formisano (University of Udine, DMIF)
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SC

arXiv:2607.21184v1 Announce Type: cross Abstract: Inductive Logic Programming (ILP) originated within the Logic Programming community in the Nineties as a framework for combining symbolic learning with declarative knowledge representation. Nowadays, mature ILP frameworks exist and they are capable o...

📖 Read original article


156. Transformer-based Diffusion models for Hydrological Time Series Probabilistic Imputation and Forecasting ​

Author: Ferdinand Bhavsar (INRAE), Lionel Benoit (INRAE), Maxime Savatier (ANDRA), Edith Gabriel (INRAE)
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.21200v1 Announce Type: cross Abstract: The modeling of hydrometeorological time series with limited observations is a key challenge in the monitoring of hydro-systems and water resources, as well as for flood or drought risk assessment. Due to the high variability of the underlying proces...

📖 Read original article


157. DART: A Degradation-Aware Recurrent Transformer for Archival Film Restoration ​

Author: Miko{\l}aj Jastrz\k{e}bski, Wojciech Koz{\l}owski, Kamil Adamczewski
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.21219v1 Announce Type: cross Abstract: Archival film restoration is a challenging problem because historical footage contains compound degradations such as scratches, dust, blur, noise, flicker, and photometric aging, while clean reference videos are unavailable. Existing video restoratio...

📖 Read original article


158. Adaptive Depth Sparse Framework: Similarity-Driven Resource Allocation for Pre-Trained LLMs ​

Author: Yidu Wu, Xiang Wang, Kejie Zhao, Zhangchi Wang, Qinghai Guo, Xiaoying Tang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.21291v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong generation and reasoning performance, but the Transformer architecture incurs high inference cost. Existing acceleration methods often rely on task-specific fine-tuning or training from scratch, increasing ...

📖 Read original article


159. Cautious optimism for deep parameterized quantum circuits ​

Author: Marie Kempkes, Elies Gil-Fuster, Carlos Bravo-Prieto, Aroosa Ijaz, Alissa Wilms, Jens Eisert, Evert van Nieuwenburg, Vedran Dunjko
Published: 7/24/2026, 4:00:00 AM
Categories: quant-ph, cs.LG, stat.ML

arXiv:2607.21409v1 Announce Type: cross Abstract: A central challenge in quantum machine learning is understanding the scaling behavior of parameterized quantum circuits (PQCs). In particular, it remains unclear how their performance on unseen data changes as the number of trainable parameters incre...

📖 Read original article


160. Token Budget Saturation and Mechanistic Early Detection of Reasoning Non-Convergence in Chain-of-Thought Models ​

Author: Renuka Oladri, Niveda Jawahar, Abdirisak Mohamed
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.21433v1 Announce Type: cross Abstract: Chain-of-thought reasoning models such as DeepSeek-R1-Distill-Qwen-7B exhibit a bimodal convergence pattern: generations either terminate within a token budget (converged) or exhaust it without reaching a conclusion (non-converged). We characterize t...

📖 Read original article


161. Climate-resilient electric vehicle charging infrastructure for sustainable cities: An interpretable causal-ensemble framework for preventive maintenance and low-carbon mobility ​

Author: Cande Lian (School of Management, Foshan University, Foshan, China), Wentao Zeng (School of Management, Foshan University, Foshan, China), Jiabin Wu (School of Management, Foshan University, Foshan, China), Yiming Bie (School of Transportation, Jilin University, Changchun, China), Wei Zhou (Department of Civil and Environmental Engineering, National University of Singapore)
Published: 7/24/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2607.21444v1 Announce Type: cross Abstract: Reliable electric vehicle (EV) charging infrastructure is a cornerstone of sustainable, low-carbon cities, yet urban climate stress such as extreme heat, heavy precipitation, and humidity increasingly raises equipment fault risk and undermines the re...

📖 Read original article


162. What, Where, and How: Disentangling the Roles of Task, Language, and Model in Code Model Representations ​

Author: Piotr Wilam
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.21491v1 Announce Type: cross Abstract: Do independently trained language models come to represent the same thing in the same way? We answer for code, extending a recently introduced concept-circuit extraction method to a 2x2 design -- Python and Rust crossed with Qwen2.5-Coder-7B and Deep...

📖 Read original article


163. Toward Generalizable Cognitive Impairment Detection with Speech-Based Multimodal Large Language Models ​

Author: Yingchao Huang, Xin Wang, Yuhan Su, Shanshan Yao
Published: 7/24/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2607.21496v1 Announce Type: cross Abstract: Cognitive impairment (CI) is a growing public health concern. Early and accurate diagnosis is critical for enabling timely intervention and improving patient outcomes. Speech-based CI detection has emerged as a promising non-invasive approach, as spe...

📖 Read original article


164. The Boundaries of Automation: A Theory of Persistent Human Participation ​

Author: Fares Fourati, Hinrich Sch"utze, Eyke H"ullermeier, Iryna Gurevych
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.ET, cs.LG, cs.MA

arXiv:2607.21547v1 Announce Type: cross Abstract: The rapid progress of AI has intensified the long-standing pursuit of automation: replacing human participation with algorithms wherever possible. Implicit in this pursuit is the assumption that humans remain in the loop only because current AI syste...

📖 Read original article


165. Neural solutions of coupled ghost and gluon Dyson--Schwinger equations in Landau gauge ​

Author: Rodrigo Carmo Terin
Published: 7/24/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-lat, hep-th

arXiv:2607.21548v1 Announce Type: cross Abstract: The coupled ghost and gluon Dyson--Schwinger equations (DSEs) of four-dimensional Landau-gauge Yang--Mills (YM) theory are solved with a neural representation trained only from renormalized equation residuals. The neural and fixed-point solutions agr...

📖 Read original article


166. MIRROR: Learning from the Other View for Multi-Modal Reasoning ​

Author: Wen Ye, Yuxiao Qu, Aviral Kumar, Xuezhe Ma
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.21552v1 Announce Type: cross Abstract: Unlike large language models (LLMs) that exhibit strong reasoning capabilities, vision-language models (VLMs) struggle with visual reasoning, even on geometry problems that admit equivalent text, diagram, and combined diagram+text views. We show that...

📖 Read original article


167. Synthetic data generation framework for quality control automation in gravure printing ​

Author: Korota Ars`ene Coulibaly, Mohamed Hamlich, Khalid Hmali, Andrea Trombin
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, eess.IV

arXiv:2607.21577v1 Announce Type: cross Abstract: Quality control in printing, particularly in rotogravure printing, still depends on slow, costly, and subjective manual inspection. Automated surface defect detection is critical for maintaining high-quality standards in rotogravure printing. Deep le...

📖 Read original article


168. Barzilai-Borwein Fails Superlinear Convergence on an Open Set of Quadratics for Every Dimension $n\geq 4$ ​

Author: Dawei Li, Xiaotian Jiang, Mingyi Hong
Published: 7/24/2026, 4:00:00 AM
Categories: math.OC, cs.AI, cs.LG

arXiv:2607.21579v1 Announce Type: cross Abstract: Barzilai--Borwein (BB) method has shown strong practical performance in continuous optimization, yet its convergence dynamics remains poorly understood. In particular, a central unresolved question is whether BB converges superlinearly for almost eve...

📖 Read original article


169. 3D-Aware VLMs with Implicit and Explicit Geometries ​

Author: Wenhao Li, Xueying Jiang, Quanhao Qian, Deli Zhao, Ran Xu, Shijian Lu, Gongjie Zhang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.21595v1 Announce Type: cross Abstract: Despite rapid progress, most existing vision-language models (VLMs) built from 2D visual inputs often struggle when handling various 3D tasks that require fine-grained spatial understanding and reasoning. To bridge this gap, we present VLM-IE3D, a un...

📖 Read original article


170. Classifier-pruned Bayesian optimization for particle accelerator tuning: Exploring temporally structured manifold of 6D beam phase space ​

Author: Mahindra Rautela, Alan Williams, Alexander Scheinker
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2412.01748v2 Announce Type: replace Abstract: Complex dynamical systems, such as particle accelerators, often require intricate and time-consuming tuning procedures to achieve optimal performance. In many cases, these procedures must also estimate the optimal system parameters governing the dy...

📖 Read original article


171. Semantics at an Angle: When Cosine Similarity Works Until It Doesn't ​

Author: Kisung You
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2504.16318v3 Announce Type: replace Abstract: Cosine similarity is a standard comparison rule for learned representations in information retrieval, natural language processing, computer vision, and multimodal learning. Its popularity is well founded: it removes positive radial scale, is comput...

📖 Read original article


172. PreMoE: Proactive Inference for Efficient Mixture-of-Experts ​

Author: Zehua Pei, Ying Zhang, Hui-Ling Zhen, Tao Yuan, Xianzhi Yu, Zhenhua Dong, Sinno Jialin Pan, Mingxuan Yuan, Bei Yu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2505.17639v4 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models offer dynamic computation, but are typically deployed as static full-capacity models, missing opportunities for deployment-specific specialization. We introduce PreMoE, a training-free framework that proactively comp...

📖 Read original article


173. Concept Concentration for Faithful Representation Intervention ​

Author: Hongzheng Yang, Yongqiang Chen, Zeyu Qin, Tongliang Liu, Chaowei Xiao, Kun Zhang, Bo Han
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2505.18672v2 Announce Type: replace Abstract: Representation intervention aims to localize and modify the representations that encode the underlying concepts in large language models (LLMs) to elicit the aligned and expected behaviors. Despite the empirical success, it has never been examined ...

📖 Read original article


174. SESaMo: Symmetry-Enforcing Stochastic Modulation for Normalizing Flows ​

Author: Janik Kreit, Dominic Schuh, Kim A. Nicoli, Lena Funcke
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.str-el, hep-lat, physics.comp-ph

arXiv:2505.19619v3 Announce Type: replace Abstract: Deep generative models have recently garnered significant attention across various fields, from physics to chemistry, where sampling from unnormalized Boltzmann-like distributions represents a fundamental challenge. In particular, autoregressive mo...

📖 Read original article


175. Tackling Heterogeneity in Federated Learning via Variance-Reduced Boltzmann Sampling within Homogeneous Social Coalitions ​

Author: Alessandro Licciardi, Roberta Raineri, Anton Proskurnikov, Lamberto Rondoni, Lorenzo Zino
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2506.02897v3 Announce Type: replace Abstract: Federated Learning (FL) enables privacy-preserving collaborative model training, but its effectiveness is often limited by client data heterogeneity. We introduce a client-selection algorithm that (i) dynamically forms nonoverlapping coalitions of ...

📖 Read original article


176. Overcoming the Communication-Performance Tradeoff in LLM Pretraining ​

Author: Amir Sarfi, Benjamin Th'erien, Joel Lidin, Eugene Belilovsky
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2508.15706v3 Announce Type: replace Abstract: Communication-efficient distributed training algorithms (e.g., DiLoCo) have received considerable interest due to their benefits for training large language models (LLMs) in bandwidth-constrained settings, such as across datacenters and over the in...

📖 Read original article


177. MO-GRPO: Mitigating Reward Hacking of Group Relative Policy Optimization on Multi-Objective Problems ​

Author: Yuki Ichihara, Yuu Jinnai, Tetsuro Morimura, Mitsuki Sakamoto, Ryota Mitsuhashi, Eiji Uchibe
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.22047v3 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has been shown to be an effective algorithm when an accurate reward model is available. However, such a highly reliable reward model is not available in many real-world tasks. In this paper, we particularly...

📖 Read original article


178. Simple Policy Gradients for Reasoning with Diffusion Language Models ​

Author: Anthony Zhan
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2510.04019v3 Announce Type: replace Abstract: Diffusion large language models (dLLMs) represent a promising alternative to autoregressive LLMs; however, the lack of effective post-training techniques, including reinforcement learning (RL), remains a key challenge for dLLMs, especially for down...

📖 Read original article


179. On the Granularity of Causal Effect Identifiability ​

Author: Yizuo Chen, Adnan Darwiche
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ME

arXiv:2510.16703v3 Announce Type: replace Abstract: The classical notion of causal effect identifiability is defined in terms of treatment and outcome variables. In this paper, we consider the identifiability of state-based causal effects: how an intervention on a particular state of treatment varia...

📖 Read original article


180. Structure-Preserving Physics-Informed Neural Network for the Korteweg--de Vries (KdV) Equation ​

Author: Victory Obieke, Emmanuel Oguadimma
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, nlin.PS, physics.flu-dyn

arXiv:2511.00418v2 Announce Type: replace Abstract: Physics-Informed Neural Networks (PINNs) offer a flexible framework for solving nonlinear partial differential equations (PDEs), yet conventional implementations often fail to preserve key physical invariants during long-term integration. This pape...

📖 Read original article


181. Black box behavioural modelling: Predicting human activity schedules with a deep conditional generative approach ​

Author: Fred Shone, Tim Hillel
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.04223v2 Announce Type: replace Abstract: Modelling the complexity and diversity of human activity scheduling behaviour is inherently challenging. We demonstrate ActVAE, a deep conditional-generative machine learning approach for the modelling of activity schedules. Suitable for applicatio...

📖 Read original article


182. CLOAK: Contrastive Guidance for Latent Diffusion-Based Data Obfuscation ​

Author: Xin Yang, Omid Ardakanian
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2512.12086v2 Announce Type: replace Abstract: Data obfuscation is a promising technique for mitigating attribute inference attacks by semi-trusted parties with access to time-series data emitted by sensors. Recent advances leverage conditional generative models together with adversarial traini...

📖 Read original article


183. Knowledge-Guided Time-Varying Causal Inference for Arctic Sea Ice Dynamics ​

Author: Akila Sampath, Vandana Janeja, Jianwu Wang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.17647v3 Announce Type: replace Abstract: Quantifying the causal relationship between sea ice thickness and sea surface height (SSH) is essential for understanding the mechanisms driving polar climate dynamics. Conventional deep learning models often struggle with treatment effect estimati...

📖 Read original article


184. NeuraLSP: A Neural Spectral Preconditioner for Accelerating PDE Solvers ​

Author: Alexander Benanti, Xi Han, Hong Qin
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.20174v3 Announce Type: replace Abstract: Solving large-scale sparse linear systems originating from partial differential equations (PDEs) is a fundamental topic in high-performance scientific computing, where preconditioners are crucial. Multigrid methods are among the most effective prec...

📖 Read original article


185. PILD: Physics-Informed Learning via Diffusion ​

Author: Tianyi Zeng, Tianyi Wang, Jiaru Zhang, Zimo Zeng, Feiyang Zhang, Yiming Xu, Sikai Chen, Junfeng Jiao, Christian Claudel, Xinbo Chen
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.ET, math.AP

arXiv:2601.21284v2 Announce Type: replace Abstract: Diffusion models have emerged as powerful generative tools for modeling complex data distributions, yet their purely data-driven nature limits applicability in engineering and scientific problems where physical laws must be respected. This paper pr...

📖 Read original article


186. Variational Speculative Decoding: Rethinking Draft Training from Token Likelihood to Sequence Acceptance ​

Author: Xiandong Zou, Jianshu Li, Jing Huang, Pan Zhou
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.PR

arXiv:2602.05774v5 Announce Type: replace Abstract: Speculative decoding accelerates inference for (M)LLMs, yet a training-decoding discrepancy persists: while existing methods optimize single greedy trajectories, decoding involves verifying and ranking multiple sampled draft paths. We propose Varia...

📖 Read original article


187. Self-Evolving Recommendation System: End-To-End Autonomous Model Optimization With LLM Agents ​

Author: Haochen Wang, Yi Wu, Daryl Chang, Li Wei, Lukasz Heldt
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.10226v2 Announce Type: replace Abstract: Optimizing large-scale machine learning systems, such as recommendation models for global video platforms, requires navigating a massive hyperparameter search space and, more critically, designing sophisticated optimizers, architectures, and reward...

📖 Read original article


188. Encode Once, Decode Never: Reusing Audio LM Internals for Efficient Temporal Localization ​

Author: Joseph An, Phillip Keung, Jiaqi Wang, Orevaoghene Ahia, Noah A. Smith
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.SD, eess.AS

arXiv:2602.10230v2 Announce Type: replace Abstract: Audio language models process input audio into rich frame-level representations, but the standard approach to temporal localization generates timestamps as sequences of text tokens, which discards the frame-level representations in favor of autoreg...

📖 Read original article


189. Wireless TokenCom: RL-Based Tokenizer Agreement for Multi-User Wireless Token Communications ​

Author: Farshad Zeinali, Mahdi Boloursaz Mashhadi, Rahim Tafazolli
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.12338v2 Announce Type: replace Abstract: Token Communications (TokenCom) has recently emerged as an effective new paradigm, where tokens are the unified units of multimodal communications and computations, enabling efficient digital semantic- and goal-oriented communications in future wir...

📖 Read original article


190. SR-TTT Does Not Learn Retrieval: A Correction and Mechanistic Post-Mortem of Surprisal-Aware Residual Test-Time Training ​

Author: Swamynathan V P
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2603.06642v2 Announce Type: replace Abstract: Test-Time Training (TTT) language models replace the KV-cache with fast weights updated during inference, achieving O(1) memory but suffering catastrophic failure on exact-recall tasks. Version 1 of this work proposed SR-TTT, which routes high-surp...

📖 Read original article


191. PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses ​

Author: Chenlong Yin, Runpeng Geng, Yanting Wang, Jinyuan Jia
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2603.13026v2 Announce Type: replace Abstract: Prompt injection poses serious security risks to real-world LLM applications, particularly autonomous agents. Although many defenses have been proposed, their robustness against adaptive attacks remains insufficiently evaluated, potentially creatin...

📖 Read original article


192. Towards Disentangled Preference Optimization Dynamics: Suppress the Loser, Preserve the Winner ​

Author: Wei Chen, Yubing Wu, Junmei Yang, Delu Zeng, Qibin Zhao, John Paisley, Min Chen, Zhou Wang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2604.18239v4 Announce Type: replace Abstract: Preference optimization is widely used to align large language models (LLMs) with human preferences. However, many margin-based methods also suppress the chosen response when they try to suppress the rejected one, and there is no general way to pre...

📖 Read original article


193. Hankel and Toeplitz Rank-1 Decomposition of Arbitrary Matrices with Applications to Signal Direction-of-Arrival Estimation ​

Author: Georgios I. Orfanidis, Dimitris A. Pados, George Sklivanitis, Elizabeth Serena Bentley
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2604.26787v2 Announce Type: replace Abstract: We consider the problems of computing the optimal rank-1 Hankel and Toeplitz-structured approximation of arbitrary matrices under L2 and L1-norm error. Such problems arise naturally in engineered systems, including the basic few-shot signal Directi...

📖 Read original article


194. Direct Bethe Free Energy Minimization for Bayesian Neural Networks ​

Author: Pavel Prochazka
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.08446v4 Announce Type: replace Abstract: Bayesian neural networks are typically trained on the evidence lower bound (ELBO), which keeps the joint likelihood but pays a Jensen gap at every observation. We train by local consistency instead: direct minimisation of the Bethe free energy, who...

📖 Read original article


195. MuonQ: Enhancing Low-Bit Muon Quantization via Directional Fidelity Optimization ​

Author: Yupeng Su, Ruijie Zhang, Ziyue Liu, Yequan Zhao, Zheng Zhang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.11396v2 Announce Type: replace Abstract: The Muon optimizer has emerged as a compelling alternative to Adam for training large language models, achieving remarkable computational savings through gradient orthogonalization. However, Muon's optimizer state is more sensitive to quantization ...

📖 Read original article


196. Understanding and Accelerating the Training of Masked Diffusion Language Models ​

Author: Chunsan Hong, Sanghyun Lee, Chieh-Hsin Lai, Satoshi Hayakawa, Yuhta Takida, Yuki Mitsufuji, Seungryong Kim, Jong Chul Ye
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2605.13026v2 Announce Type: replace Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models (ARMs) for language modeling. However, MDMs are known to learn substantially more slowly than ARMs, which may become problematic when scaling MDMs to la...

📖 Read original article


197. Constrained latent state modeling: A unifying perspective on representation learning under competing constraints ​

Author: Gwenol'e Quellec
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.15995v2 Announce Type: replace Abstract: Learning latent representations from complex data is central to modern machine learning, spanning temporal, multimodal, and partially observed systems. In such settings, representations are more naturally understood as latent states capturing under...

📖 Read original article


198. Dimension-Free Convergence of Discrete Diffusion Models: Adjoint Equations Induce the Right Space ​

Author: Kelvin Kan, Xingjian Li, Benjamin J. Zhang, Tuhin Sahai, Stanley Osher, Markos A. Katsoulakis
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.ML, stat.TH

arXiv:2605.17232v4 Announce Type: replace Abstract: Discrete diffusion has become a leading framework for generative modeling in various applications including language, vision, and biology. Existing convergence theory, however, exhibits fundamental limitations. KL-based analyses diverge under singu...

📖 Read original article


199. Learning Transfers: Kan Extensions for Neural Invariants ​

Author: Luciano Melodia
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, math.AT, math.CT

arXiv:2606.07627v2 Announce Type: replace Abstract: Transfer learning presumes that a representation learned on a source task carries structure that remains usable on a related target task. Standard evaluations probe this through target accuracy or a distributional discrepancy, without stating which...

📖 Read original article


200. Tight Sample Complexity of Transformers ​

Author: Chenxiao Yang, Nathan Srebro, Zhiyuan Li
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.09731v2 Announce Type: replace Abstract: We tightly characterize the VC dimension of depth-$L$ Transformers with a total of $W$ parameters, mapping an input sequence of length $T$ to a single output, establishing an upper bound of $O(L W \log (T W))$ and a nearly matching lower bound of $...

📖 Read original article


201. SymQNet: Amortized Acquisition for Low-Latency Adaptive Hamiltonian Learning ​

Author: Yash Vardhan Tomar, Dheeraj Peddireddy
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.12808v4 Announce Type: replace Abstract: Adaptive Hamiltonian learning is central to calibrating and characterizing quantum devices. In an adaptive controller, choosing the next experiment is itself a computation. Bayesian design rules are recomputed after every posterior update, and that...

📖 Read original article


202. HAMON: Passive Optical Sequence Mixing for Long-Horizon Forecasting ​

Author: Alper Y{\i}ld{\i}r{\i}m
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.AR

arXiv:2606.17028v2 Announce Type: replace Abstract: Simple linear and frequency-domain models remain surprisingly competitive in long-horizon time-series forecasting, and recent mechanistic evidence suggests that standard forecasting benchmarks may not require the dense superposed representations th...

📖 Read original article


203. Hierarchical Self-Supervised Representation Learning Framework for Multivariate Time Series Grounded in ECG Analysis ​

Author: Siwon Kim
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2607.01145v3 Announce Type: replace Abstract: Data analysis in the medical domain often encounters scenarios involving a limited target dataset and a large, unannotated dataset with a general distribution. Under such circumstances, self-supervised learning (SSL) methods are highly effective fo...

📖 Read original article


204. Parameter-Free Encoders Remain Viable for RDB Foundation Models ​

Author: Linjie Xu, David Wipf
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.05476v2 Announce Type: replace Abstract: Given a relational database (RDB) storing heterogeneous tabular information, how can we predict missing (or future) values in some target column of interest? As the space of potential targets is vast across enterprise settings, it is preferable to ...

📖 Read original article


205. TabPFN beyond Tabular Data: Calibration and Accuracy on Multimodal Embeddings ​

Author: Jingxiang Zhang, Lujia Zhong, Zijie Zhu, Shuo Huang, Yuang Xu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11007v2 Announce Type: replace Abstract: Few-shot multimodal classification commonly attaches a lightweight head, such as $k$-nearest neighbors, logistic regression, or a linear SVM, to a frozen pretrained encoder. Although computationally efficient, these heads can produce poorly calibra...

📖 Read original article


206. RF Spectrogram Anomaly Detection with Quantum Kitchen Sinks: Architecture, Representation, and Hardware Validation ​

Author: Abdallah Aaraba, Alexis Vieloszynski, Remon Polus, Ola Ahmad, Soumaya Cherkaoui
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.13897v2 Announce Type: replace Abstract: The broadcast nature of wireless channels exposes radio-frequency (RF) networks to anomalous and malicious transmissions, making anomaly detection a fundamental requirement for secure spectrum management. Quantum Kitchen Sinks (QKS) offer a lightwe...

📖 Read original article


207. FlashPDE: A Drop-In Fused Triton Operator Library for Neural PDE Solvers ​

Author: Peiyu Zang, Bosen Xie, Ruoxiang Xu, Yongqiang Cai
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.MS

arXiv:2607.18020v2 Announce Type: replace Abstract: Physics-Informed Neural Networks (PINNs) solve PDEs by incorporating physical constraints into neural-network training, but large-scale problems are limited by automatic-differentiation memory overhead and inefficient execution of grid-based PDE op...

📖 Read original article


208. A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space ​

Author: Shuangyao Huang
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.MA

arXiv:2607.18597v2 Announce Type: replace Abstract: Counterfactual credit assignment has proven effective in multi-agent reinforcement learning (MARL) for discrete action spaces, yet its extension to continuous-action cooperative tasks remains challenging. Existing methods that approximate the count...

📖 Read original article


209. Probabilistic Physics-Aware Machine Learning Predictions of Electric Truck Energy Consumption with Field Data ​

Author: Hannes Nilsson, Rafael Basso, Bal'azs Kulcs'ar, Morteza Haghir Chehreghani
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19054v2 Announce Type: replace Abstract: In this work, we incorporate first principle physics into the construction of data-driven methods by considering a model that accounts for the different sources of energy losses during vehicle operations. Our results show that Bayesian linear regre...

📖 Read original article


210. Riemannian Deep Learning: Modules, Networks, and Geometries ​

Author: Chen Ziheng
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.DG

arXiv:2607.19305v2 Announce Type: replace Abstract: Deep neural networks on manifold-valued representations have attracted growing interest, but many basic components remain tied to specific manifolds, rely on Euclidean approximations, or require costly and numerically fragile geometric operations. ...

📖 Read original article


211. Neural Operator Surrogates for Two-Dimensional Neutron Flux Estimation ​

Author: Japan K. Patel, Barry D. Ganapol, Anthony Magliari, Matthew C. Schmidt, Todd A. Wareing
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19388v2 Announce Type: replace Abstract: This work extends our one-dimensional single-sweep neural-operator studies to two dimensions. We consider one-group transport with isotropic scattering. As in the one-dimensional work, we use Fourier neural operators (FNOs) to approximate the high-...

📖 Read original article


212. The Orthogonalized Read Is a Removable Training Scaffold for Recurrent Memory ​

Author: Keston Aquino-Michaels
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19390v2 Announce Type: replace Abstract: A recent report finds that orthogonalizing the mLSTM memory matrix at read time (five Newton-Schulz iterations, trained through) substantially improves noisy associative recall. The effect replicates, but it is not a memory improvement. Training on...

📖 Read original article


213. The C-index illusion: discrimination without calibration in published survival models ​

Author: Rafael da Silva, Danilo Alvares
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.19526v2 Announce Type: replace Abstract: Recent work has argued normatively, on synthetic data, that evaluating survival models by discrimination alone (concordance index) yields systematically misleading model comparisons, because the metric ignores calibration and time-dependent accurac...

📖 Read original article


214. OPIUM: Mitigating Steering Externalities and Over-Refusal via Dual Objective Latent Optimization ​

Author: Kavin Aravindan, Arihant Rastogi, Krishak Aneja, Aadi Prasad, Saiyam Jain, Vaishnavi Shivkumar, Ponnurangam Kumaraguru
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19806v2 Announce Type: replace Abstract: Activation steering provides a lightweight mechanism for controlling large language models at inference time, but steering vectors can have unintended externalities: utility vectors may weaken safety behavior, while refusal vectors may induce over-...

📖 Read original article


215. Generalized Kalman filter based temporal difference reinforcement learning ​

Author: Vasos Arnaoutis, Eric Lutters, Bojana Rosi'c
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.CE

arXiv:2607.20010v2 Announce Type: replace Abstract: In this paper, we present a generalized temporal-difference (TD) reinforcement learning framework based on the theory of conditional expectations. The value and action-value (Q-value) functions are treated as uncertain quantities, and their estimat...

📖 Read original article


216. Co-Evolving LLM Evaluators and Policies via DynamicRubric ​

Author: Beining Wang, Weihang Su, Hongtao Tian, Hao Kong, Tao Yang, Ting Yao, Qingyi Pan, Yueyue Wu, Qingyao Ai, Min Zhang, Yiqun Liu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.20083v2 Announce Type: replace Abstract: Post-training with evaluator feedback on policy-induced samples serves as a major mechanism for improving large language models. As policies improve, these sampled responses become close in quality. These close candidates create a bottleneck for po...

📖 Read original article


217. Instance Hardness-Based Relevance for Imbalanced Regression ​

Author: Vitor M. Leitao, Juscimara G. Avelino, George D. C. Cavalcanti, Rafael M. O. Cruz
Published: 7/24/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.20173v2 Announce Type: replace Abstract: Imbalanced regression problems arise when the target variable has an asymmetric distribution, resulting in underrepresented value ranges in the dataset. Traditional approaches for identifying rare instances rely on a relevance function that assigns...

📖 Read original article


218. Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning ​

Author: Jia Wan, Sean R. Sinclair, Devavrat Shah, Martin J. Wainwright
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC

arXiv:2409.14557v4 Announce Type: replace-cross Abstract: We study a structured class of Markov Decision Processes, known as Exo-MDPs, in which the state space is partitioned into exogenous and endogenous components. Exogenous states evolve stochastically, independent of the agent's actions, while e...

📖 Read original article


219. Evaluation and Prognostic Validation of Deep Regression Models for WSI-Based Gene-Expression Prediction ​

Author: Fredrik K. Gustafsson, Constance Boissin, Johan Vallon-Christersson, Mattias Rantalainen
Published: 7/24/2026, 4:00:00 AM
Categories: q-bio.GN, cs.CV, cs.LG

arXiv:2410.00945v2 Announce Type: replace-cross Abstract: Gene-expression profiling is widely used in research and central to many areas of precision oncology, but remains costly and not universally accessible. Recent advances in computational pathology enable prediction of transcriptomic profiles d...

📖 Read original article


220. Deepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook ​

Author: Florinel-Alin Croitoru, Andrei-Iulian Hiji, Vlad Hondru, Nicolae Catalin Ristea, Paul Irofti, Marius Popescu, Cristian Rusu, Radu Tudor Ionescu, Fahad Shahbaz Khan, Mubarak Shah
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, cs.MM, cs.SD, eess.AS

arXiv:2411.19537v3 Announce Type: replace-cross Abstract: We survey deepfake generation and detection techniques, covering all deepfake media types: image, video, audio and multimodal content. We identify various kinds of deepfakes and construct taxonomies of deepfake generation and detection method...

📖 Read original article


221. How Robust Is Homogeneity Bias in LLMs? Evidence Across Models, Decoding Settings, and Identity Signals ​

Author: Messi H. J. Lee
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2501.02211v3 Announce Type: replace-cross Abstract: Large language models (LLMs) reproduce homogeneity bias -- the tendency to portray marginalized groups as more internally similar than dominant groups -- but whether this bias generalizes across models, is stable under different inference set...

📖 Read original article


222. Statistical Inference for Generative Model Comparison ​

Author: Zijun Gao, Yan Sun, Han Su
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2501.18897v4 Announce Type: replace-cross Abstract: Generative models have achieved remarkable success across a range of applications, yet their evaluation still lacks principled uncertainty quantification. In this paper, we develop a method for comparing how close different generative models ...

📖 Read original article


223. Enhancing Membership Inference Attacks on Diffusion Models from a Frequency-Domain Perspective ​

Author: Puwei Lian, Yujun Cai, Songze Li, Bingkun Bao
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2505.20955v5 Announce Type: replace-cross Abstract: Diffusion models have achieved tremendous success in image generation, but they also raise significant concerns regarding privacy and copyright issues. Membership Inference Attacks (MIAs) are designed to ascertain whether specific data was ut...

📖 Read original article


224. Gibbs randomness-compression proposition ​

Author: M S"uzen
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2505.23869v5 Announce Type: replace-cross Abstract: A proposition that connects randomness and compression is put forward via Gibbs entropy over set of measurement vectors associated with a lossy compression process. In building this connection, we use a performance of a learning task as a pro...

📖 Read original article


225. Loss-Complexity Landscape and Model Structure Functions ​

Author: Alexander Kolpakov
Published: 7/24/2026, 4:00:00 AM
Categories: cs.IT, cs.AI, cs.LG, math-ph, math.IT, math.MP

arXiv:2507.13543v5 Announce Type: replace-cross Abstract: We develop a framework for dualizing the Kolmogorov structure function $h_x(\alpha)$, which then allows using computable complexity proxies. We establish a mathematical analogy between information-theoretic constructs and statistical mechanic...

📖 Read original article


226. DynaMark: A Reinforcement Learning Framework for Dynamic Watermarking in Industrial Machine Tool Controllers ​

Author: Navid Aftabi, Abhishek Hanchate, Satish Bukkapatnam, Dan Li
Published: 7/24/2026, 4:00:00 AM
Categories: eess.SY, cs.AI, cs.CR, cs.LG, cs.SY, stat.AP

arXiv:2508.21797v2 Announce Type: replace-cross Abstract: Industry 4.0's highly networked Machine Tool Controllers (MTCs) are prime targets for replay attacks that use outdated sensor data to manipulate actuators. Dynamic watermarking can reveal such tampering, but current schemes assume linear-Gaus...

📖 Read original article


227. torchsom: The Reference PyTorch Library for Self-Organizing Maps ​

Author: Louis Berthier, Ahmed Shokry, Maxime Moreaud, Guillaume Ramelet, Eric Moulines
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2510.11147v2 Announce Type: replace-cross Abstract: This paper introduces torchsom, an open-source Python library that provides a reference implementation of the Self-Organizing Map (SOM) in PyTorch. This package offers three main features: (i) dimensionality reduction, (ii) clustering, and (i...

📖 Read original article


228. Neural Guided Sampling for Quantum Circuit Optimization ​

Author: Bodo Rosenhahn, Tobias J. Osborne, Christoph Hirche
Published: 7/24/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2510.12430v2 Announce Type: replace-cross Abstract: Translating a general quantum circuit on a specific hardware topology with a reduced set of available gates, also known as transpilation, comes with a substantial increase in the length of the equivalent circuit. Due to decoherence, the quali...

📖 Read original article


229. Statistical physics of deep learning: Optimal learning of a multi-layer perceptron near interpolation ​

Author: Jean Barbier, Francesco Camilli, Minh-Toan Nguyen, Mauro Pastore, Rudy Skerk
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.dis-nn, cond-mat.stat-mech, cs.IT, cs.LG, math.IT

arXiv:2510.24616v4 Announce Type: replace-cross Abstract: For four decades statistical physics has been providing a framework to analyse neural networks. A long-standing question remained on its capacity to tackle deep learning models capturing rich feature learning effects, thus going beyond the na...

📖 Read original article


230. Spectral functions in Minkowski quantum electrodynamics from neural reconstruction ​

Author: Rodrigo Carmo Terin
Published: 7/24/2026, 4:00:00 AM
Categories: hep-ph, cs.LG, hep-lat, hep-th

arXiv:2510.24728v2 Announce Type: replace-cross Abstract: We study neural reconstructions of quenched rainbow quantum electrodynamics (QED) Dyson--Schwinger benchmarks in Minkowski-related kinematics. Using the dispersive formulation as motivation, we separate the Euclidean Fukuda--Kugo equation, th...

📖 Read original article


231. Limit Theorems for Stochastic Gradient Descent in High-Dimensional Single-Layer Networks ​

Author: Parsa Rangriz
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR, math.ST, stat.TH

arXiv:2511.02258v3 Announce Type: replace-cross Abstract: This paper studies the high-dimensional scaling limits of online stochastic gradient descent (SGD). Building on the work of Ben Arous, Gheissari, and Jagannath on the effective dynamics of SGD, we study the critical scaling regime of the step...

📖 Read original article


232. Compiling to recurrent neurons ​

Author: Joey Velez-Ginorio, Nada Amin, Konrad Kording, Steve Zdancewic
Published: 7/24/2026, 4:00:00 AM
Categories: cs.PL, cs.LG

arXiv:2511.14953v2 Announce Type: replace-cross Abstract: Discrete structures are currently second-class in differentiable programming. Since functions over discrete structures lack overt derivatives, differentiable programs do not differentiate through them and limit where they can be used. For exa...

📖 Read original article


233. Environment-Aware Channel Inference via Cross-Modal Flow: From Multimodal Sensing to Wireless Channels ​

Author: Guangming Liang, Mingjie Yang, Dongzhu Liu, Paul Henderson, Lajos Hanzo
Published: 7/24/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, eess.SP, math.IT

arXiv:2512.04966v2 Announce Type: replace-cross Abstract: Accurate channel state information (CSI) underpins reliable and efficient wireless communication. However, acquiring CSI via pilot estimation incurs substantial overhead, especially in massive multiple-input multiple-output (MIMO) systems ope...

📖 Read original article


234. Exploiting the Alternatives: Coordinated Learning via Hierarchical RL for Dynamic VNEAP ​

Author: Ali Al Housseini, Cristina Rottondi, Sebastian Troia, Omran Ayoub
Published: 7/24/2026, 4:00:00 AM
Categories: cs.NI, cs.LG, cs.MA

arXiv:2512.05207v3 Announce Type: replace-cross Abstract: Virtual Network Embedding (VNE) is a key enabler of network slicing, yet most formulations assume that each Virtual Network Request (VNR) has a fixed topology. Recently, VNE with Alternatives (VNEAP) was introduced to capture malleable VNRs, ...

📖 Read original article


235. Minimum Bayes Risk Decoding for Error Span Detection in Reference-Free Automatic Machine Translation Evaluation ​

Author: Boxuan Lyu, Haiyue Song, Hidetaka Kamigaito, Chenchen Ding, Hideki Tanaka, Masao Utiyama, Kotaro Funakoshi, Manabu Okumura
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2512.07540v4 Announce Type: replace-cross Abstract: Error Span Detection (ESD) extends automatic machine translation (MT) evaluation by localizing translation errors and labeling their severity. Current generative ESD methods typically use Maximum a Posteriori (MAP) decoding, assuming that the...

📖 Read original article


236. Interpretable Embeddings with Sparse Autoencoders: A Data Analysis Toolkit ​

Author: Nick Jiang, Xiaoqing Sun, Lisa Dunlap, Lewis Smith, Neel Nanda
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2512.10092v2 Announce Type: replace-cross Abstract: Analyzing large-scale text corpora is a core challenge in machine learning, crucial for tasks like identifying undesirable model behaviors or biases in training data. Current methods often rely on costly LLM-based techniques (e.g. annotating ...

📖 Read original article


237. Boltzmann generators for amorphous particle systems ​

Author: Louis Grenioux, Leonardo Galliano, Ludovic Berthier, Giulio Biroli, Marylou Gabri'e
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cond-mat.stat-mech, cs.LG, physics.comp-ph

arXiv:2512.16607v2 Announce Type: replace-cross Abstract: Sampling configurations in thermodynamic equilibrium is a long-standing challenge in statistical physics. Boltzmann generators address this problem by employing generative models to propose independent configurations, which are then reweighte...

📖 Read original article


238. Non-Stationary Functional Bilevel Optimization ​

Author: Jason Bohne, Ieva Petrulionyte, Michael Arbel, Julien Mairal, Pawe{\l} Polak
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2601.15363v2 Announce Type: replace-cross Abstract: Functional bilevel optimization (FBO) provides a powerful framework for hierarchical learning in function spaces, yet current methods are limited to static offline settings and perform suboptimally in online, non-stationary scenarios. We prop...

📖 Read original article


239. Error Amplification Limits ANN-to-SNN Conversion in Continuous Control ​

Author: Zijie Xu, Zihan Huang, Yiting Dong, Kang Chen, Wenxuan Liu, Zhaofei Yu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2601.21778v3 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs) can achieve competitive performance by converting already existing well-trained Artificial Neural Networks (ANNs), avoiding further costly training. This property is particularly attractive in Reinforcement Lear...

📖 Read original article


240. Multimodal Learning for Arcing Detection in Pantograph-Catenary Systems ​

Author: Hao Dong, Eleni Chatzi, Olga Fink
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2602.08792v2 Announce Type: replace-cross Abstract: The pantograph-catenary interface is essential for ensuring uninterrupted and reliable power delivery in electrified rail systems. However, electrical arcing at this interface poses serious risks, including accelerated wear of contact compone...

📖 Read original article


241. TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics ​

Author: Shirui Chen, Cole Harrison, Ying-Chun Lee, Angela Jin Yang, Zhongzheng Ren, Lillian J. Ratliff, Jiafei Duan, Dieter Fox, Ranjay Krishna
Published: 7/24/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2602.19313v2 Announce Type: replace-cross Abstract: General-purpose robot learning requires dense, instruction-conditioned feedback that can distinguish meaningful task progress from stalled, failed, or partially completed behavior. Yet obtaining such feedback at scale remains difficult, since...

📖 Read original article


242. AG-REPA: Causal Layer Selection for Representation Alignment in Audio Flow Matching ​

Author: Pengfei Zhang, Tianxin Xie, Minghao Yang, Li Liu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG, cs.MM

arXiv:2603.01006v3 Announce Type: replace-cross Abstract: REPresentation Alignment (REPA) improves the training of generative flow models by aligning intermediate hidden states with pretrained teacher features, but its effectiveness in token-conditioned audio Flow Matching critically depends on the ...

📖 Read original article


243. VPWEM: Non-Markovian Visuomotor Policy with Working and Episodic Memory ​

Author: Yuheng Lei, Zhixuan Liang, Hongyuan Zhang, Ping Luo
Published: 7/24/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2603.04910v2 Announce Type: replace-cross Abstract: Imitation learning from human demonstrations has achieved significant success in robotic control, yet most visuomotor policies still condition on single-step observations or short-context histories, making them struggle with non-Markovian tas...

📖 Read original article


244. Gumbel Distillation for Parallel Text Generation ​

Author: Chi Zhang, Xixi Hu, Bo Liu, Qiang Liu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2603.22216v2 Announce Type: replace-cross Abstract: The slow, sequential nature of autoregressive (AR) language models has driven the adoption of parallel decoding methods. However, these non-AR models often sacrifice generation quality as they struggle to model the complex joint distribution ...

📖 Read original article


245. Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of Mind ​

Author: Hanqi Xiao, Vaidehi Patil, Zaid Khan, Hyunji Lee, Elias Stengel-Eskin, Mohit Bansal
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.11666v2 Announce Type: replace-cross Abstract: As large language models (LLMs) become the engine behind conversational systems, their ability to reason about the intentions and states of their dialogue partners (i.e., form and use a theory-of-mind, or ToM) becomes increasingly critical fo...

📖 Read original article


246. Fairness Constraints in High-Dimensional Generalized Linear Models ​

Author: Yixiao Lin, James Booth
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2604.16610v3 Announce Type: replace-cross Abstract: Most fairness-aware learning methods assume that sensitive attributes are observed, an assumption that may fail due to privacy, legal, or data-collection constraints. We develop a framework for fairness-aware generalized linear models when th...

📖 Read original article


247. AI Security Policy Should Assess Systems, Not Only Models ​

Author: Michael A. Riegler, Inga Str"umke
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2605.09504v2 Announce Type: replace-cross Abstract: We present swarm-attack, an open-source adversarial testing framework in which multiple lightweight LLM agents coordinate through shared memory, parallel exploration, and evolutionary optimization. Together, our results demonstrate that both ...

📖 Read original article


248. Conditional Entropy of Heat Diffusion on Temporal Networks ​

Author: Samuel Koovely, Alexandre Bovet
Published: 7/24/2026, 4:00:00 AM
Categories: cs.SI, cond-mat.stat-mech, cs.IT, cs.LG, math.IT, physics.data-an

arXiv:2605.21514v2 Announce Type: replace-cross Abstract: Diffusion-based information-theoretic approaches provide new theoretical and practical tools to study complex networks. So far, they have not been generalized to temporal networks. In this work, we show that common entropic measures based on ...

📖 Read original article


249. Multimodality Stacking with Blockwise missing values and application to the PIONeeR biomarkers study for prediction of resistance to immunotherapy ​

Author: Mohamed Boussena, Florence Monville, Jacques Fieschi-Meric, Frederic Vely, Pierre Milpied, Julien Mazieres, Maurice Perol, Eric Vivier, Laurent Greillier, Fabrice Barlesi, Sebastien Benzekry
Published: 7/24/2026, 4:00:00 AM
Categories: stat.AP, cs.LG, q-bio.QM, stat.ML

arXiv:2605.25050v2 Announce Type: replace-cross Abstract: Integrating multimodal datasets in clinical oncology is frequently hindered by high dimensionality and blockwise missingness, where entire data sources are unavailable for specific patient subsets. Standard survival models often struggle with...

📖 Read original article


250. Theoretical Aspects of Lie Groupoid and Lie Algebroid Equivariant Convolutional Neural Networks ​

Author: Michael Astwood
Published: 7/24/2026, 4:00:00 AM
Categories: math.DG, cs.LG, math.CT

arXiv:2606.02758v2 Announce Type: replace-cross Abstract: We introduce Lie groupoid equivariant neural networks as a specialization of recently proposed topological category-equivariant neural networks to the differentiable setting. Lie groupoid equivariant neural networks are composed from Lie grou...

📖 Read original article


251. OLIVE: Online Low-Rank Incremental Learning for Efficient Adaptive Exoskeletons ​

Author: Dong Liu, Yanxuan Yu, Ben Lengerich, Tong Geng, Ying Nian Wu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2606.05234v2 Announce Type: replace-cross Abstract: Wearable exoskeleton systems hold promise for restoring mobility in individuals with physical impairments, yet most existing controllers rely on static gait policies that cannot adapt to dynamic real-world environments or individual user char...

📖 Read original article


252. Graph Reinforcement Learning for Calibration-Aware Quantum Circuit Routing ​

Author: Yash Vardhan Tomar, Dheeraj Peddireddy
Published: 7/24/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG

arXiv:2606.12816v4 Announce Type: replace-cross Abstract: Quantum circuit routing is a key step in compiling programs for noisy intermediate-scale quantum processors, particularly superconducting devices whose sparse fixed coupling makes routing a central compilation cost. Routes that appear efficie...

📖 Read original article


253. Zero-Shot Neural Priors for Generalizable Cross-Subject and Cross-Task EEG Decoding ​

Author: Baimam Boukar Jean Jacques, Brandone Fonya, Nchofon Tagha Ghogomu, Pauline Nyaboe, Kipngeno Koech
Published: 7/24/2026, 4:00:00 AM
Categories: eess.SP, cs.HC, cs.LG

arXiv:2606.23706v2 Announce Type: replace-cross Abstract: The development of generalizable electroencephalography (EEG) decoding models is essential for robust brain-computer interfaces (BCI) and objective neural biomarkers in mental health. Conventional approaches have been hindered by poor cross-s...

📖 Read original article


254. On Pairwise Quantile Regression - Statistical Guarantees and Applications ​

Author: Romain Th'er'ezien, Stephan Cl'emen\c{c}on, Fantin Girard, Hamza El-Abdouni
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.CV, cs.LG

arXiv:2607.04431v2 Announce Type: replace-cross Abstract: Quantile regression provides a powerful tool for summarizing the conditional distribution of a real-valued random variable (r.v.) of interest $Y$ as a function of covariates $Z$ in cases where it shows a large dispersion with high probability...

📖 Read original article


255. Heat-Kernel Entropy Profiles and Geometric Effective Sample Size for Weighted Measures on Manifolds ​

Author: Kisung You
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2607.06696v2 Announce Type: replace-cross Abstract: Weighted empirical measures on compact manifolds appear in importance sampling, particle approximations, posterior summaries, quadrature, and representation learning. Ordinary effective sample size and related weight summaries ignore the geom...

📖 Read original article


256. A Sovereign, Open-Source Foundation Model for German and English ​

Author: Soofi-Team, :, Benedikt Droste, David Fitzek, Ruben H"arle, Lukas Helff, Maximilian Idahl, Alex Jude, Abbas Goher Khan, Maurice Kraus, Timm Ruland, Richard Rutmann, Sebastian Sztwiertnia, Markus Frey, Daniil Gurgurov, Jan Pfister, Tom R"ohr, Sebastian von Rohrscheidt, J"org Bienert, Nicolas Flores-Herr, Simon Gottschalk, Andreas Hotho, Kristian Kersting, Joachim K"ohler, Alexander L"oser, Wolfgang Nejdl, Simon Ostermann, Jan Plogsties, Bj"orn Pl"uster, Patrick Putzky, Mehdi Ali, Michael Fromm, Max L"ubbering
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2607.09424v3 Announce Type: replace-cross Abstract: We present Soofi S 30B-A3B, a sovereign, open-source Mixture-of-Experts (MoE) hybrid Mamba Transformer foundation model for German and English. Its hybrid design activates only 3B of 30B parameters per token and keeps the inference cache near...

📖 Read original article


257. Silent Failures in Quantized LLM Reasoning: A Taxonomy-Based Analysis of Hollow Convergence and Failure Mode Shifts ​

Author: Renuka Oladri, Mohan Vamsi Varadaraju Priya, Jerry Wu
Published: 7/24/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.09999v2 Announce Type: replace-cross Abstract: We show that post-training quantization can silently alter how large language models reason even when task accuracy is preserved. Using a six-category failure taxonomy validated by two independent human annotators (Cohen's $\kappa$ = 0.906), ...

📖 Read original article


258. Spectral Concentration and Recovery in Sparse High-Dimensional Random Geometric Graphs ​

Author: Manuel Fernandez V, Yizhe Zhu
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR, math.ST, stat.TH

arXiv:2607.14304v2 Announce Type: replace-cross Abstract: We study sparse threshold random geometric graphs generated by high-dimensional spherical or Gaussian latent vectors. Although each edge has marginal probability $p$, shared latent variables make the adjacency entries dependent. At the connec...

📖 Read original article


259. Can AI Agents Really Complete RTL-to-GDS? Lessons from Benchmarking Tool-Interactive EDA Workflows ​

Author: Jinyuan Deng, Zhengrui Chen, Xufeng Wei, Tianyu Xing, Chenyi Wen, Qi Sun, Cheng Zhuo
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.AR, cs.LG

arXiv:2607.17528v3 Announce Type: replace-cross Abstract: Large language model (LLM) agents are extending electronic design automation (EDA) beyond static RTL generation toward long-horizon, tool-interactive workflows. Yet it remains unclear whether general-purpose coding agents, even with domain-sp...

📖 Read original article


260. Wisdom of LLM Crowds: Aggregation and Contamination in Language Model Ensembles ​

Author: Igor Douven
Published: 7/24/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.18269v2 Announce Type: replace-cross Abstract: The wisdom of crowds -- the finding that aggregating judgments across individuals often outperforms the best individual -- has been extensively studied with human forecasters. Whether the same phenomenon emerges when the ``crowd'' consists of...

📖 Read original article


261. The Price of Hidden Curvature: An $\widetilde{\Omega} (d^{5/4} \sqrt{T})$ Lower Bound for Bandit Convex Optimization ​

Author: Nived Rajaraman
Published: 7/24/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT

arXiv:2607.18652v2 Announce Type: replace-cross Abstract: We establish a $\widetilde\Omega(d^{5/4}\sqrt T)$ lower bound on the minimax expected regret of stochastic bandit convex optimization of $1$-Lipschitz functions on the Euclidean ball. This presents the first nontrivial regret lower bound that...

📖 Read original article


262. Multi-stage Dynamic Selection for Cross-Project Defect Prediction ​

Author: Juscimara G. Avelino, Juscelino S. A. Junior, George D. C. Cavalcanti, Rafael M. O. Cruz
Published: 7/24/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2607.20151v2 Announce Type: replace-cross Abstract: Cross-Project Defect Prediction (CPDP) involves building models using data from external projects, called training projects, to predict modules from the target project. However, traditional CPDP methods suffer from the distribution shift betw...

📖 Read original article


263. Label-Free Finite-Volume-Residual Training of Attention Graph Neural Networks for Coupled Thermo-Fluid Fields ​

Author: Tianyu Li, Zhiwei Cao, Qingang Zhang, Ruihang Wang, Binyang Song, Yonggang Wen
Published: 7/24/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG

arXiv:2607.20321v2 Announce Type: replace-cross Abstract: Neural surrogates are widely used in scientific machine learning for fast prediction of three-dimensional (3D) thermo-fluid fields. However, generating training data using conventional numerical solvers often incurs substantial computational ...

📖 Read original article