Skip to content

arXiv cs.LG - 2026-07-15 ​

234 items collected.


1. OmniPMNet: Bridging discrete and gridded PM10 forecasts via omni-query neural processes ​

Author: Shuangshuang He, Shuo Wang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, physics.ao-ph

arXiv:2607.11896v1 Announce Type: new Abstract: Forecasting particulate matter (PM10) requires both station-scale accuracy and continuous spatial fields, especially during severe dust storms. Chemical transport models (CTMs) provide gridded forecasts but retain local biases, whereas graph neural net...

📖 Read original article


2. Semidirect Fourier Delta Attention: Phase-Controlled Delta Memory with Constructive Chunk-WY Kernels ​

Author: Tiantian Zhang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11897v1 Announce Type: new Abstract: Linear attention replaces softmax attention's growing KV cache with a fixed recurrent state, but this compression limits exact state tracking and long-context memory. We introduce \emph{Semidirect Fourier Delta Attention} (SFDA), a phase-controlled gen...

📖 Read original article


3. Repairing Shape-Prior Shortcuts in Long-Range Single-Shot Fringe Projection Profilometry ​

Author: Adam Haroon, Cody Fleming, Beiwen Li
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, eess.IV

arXiv:2607.11928v1 Announce Type: new Abstract: Single-shot fringe projection profilometry (FPP) networks that regress depth directly can exploit a shape-prior shortcut, recovering depth from object boundaries rather than from fringe phase. On a photorealistic synthetic benchmark (15,600 fringe imag...

📖 Read original article


4. Qubit-Efficient Quantum Search for Hyperdimensional Decomposition via Logarithmic Encoding ​

Author: Sanggeon Yun, Hyunwoo Oh, Ryozo Masukawa, Raheeb Hassan, Mohsen Imani
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11936v1 Announce Type: new Abstract: Hyperdimensional Computing (HDC) represents symbols using high-dimensional hypervectors of dimension $D$. In hypervector decomposition, the objective is to recover $F$ constituent hypervectors, each drawn from a codebook of size $N$, from a bound targe...

📖 Read original article


5. Mirror Horizon: Viable Path Entropy as a Measure of Bounded Reflection ​

Author: Tiantian Zhang (Crystal)
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11937v1 Announce Type: new Abstract: Mirror Theory proposes that an intelligent system should be studied not only by what it represents, but by what coherent continuations it can sustain under repeated reflection. We make this claim operational through \emph{viable path entropy} (VPE), a ...

📖 Read original article


6. Mathematics of Data Science ​

Author: Afonso S. Bandeira, Amit Singer, Thomas Strohmer
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT, math.PR

arXiv:2607.11938v1 Announce Type: new Abstract: This book is about the mathematical foundations of data science. 1. Introduction 2. Curses, Blessings, and Surprises in High Dimensions 3. Singular Value Decomposition and Principal Component Analysis 4. Linear Regression and Regularization 5. Graphs, ...

📖 Read original article


7. CARE-LoRA: Compressed Activation REconstruction for Memory-Efficient LoRA ​

Author: Gengyu Zhang, Haiyin Ran, Zhengbao He, Yuhang Liu, Hanling Tian, Zhehao Huang, Xiaolin Huang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11940v1 Announce Type: new Abstract: As the scale of large pre-trained models continues to grow, fine-tuning them under limited memory budgets has become increasingly challenging. Low-Rank Adaptation (LoRA), currently one of the most widely adopted parameter-efficient fine-tuning (PEFT) m...

📖 Read original article


8. How Query Visibility Changes KV-Cache Compression Rankings: A Matched-Budget Audit ​

Author: Daming Luo, Christy Liang, Junyu Xuan
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11942v1 Announce Type: new Abstract: KV-cache compression methods are predominantly evaluated with the query appended to the context before compression -- a query-aware protocol. Yet the economic case for a compressed KV cache is reuse: compress a document once, answer many future questio...

📖 Read original article


9. BattVAE-GP: Generative Modeling of Long-Horizon Battery Degradation with Uncertainty Quantification ​

Author: Raghvender Raghvender, Mahdi Abid, Ferran Brosa Planella, Charles Delacourt, Arnaud Demorti`ere
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11943v1 Announce Type: new Abstract: Long-horizon physics-based simulations of battery degradation provide mechanistic insight but remain computationally expensive, limiting their use for dense exploration of operating conditions over extended cycle life. Here, we propose a hybrid physics...

📖 Read original article


10. Generalized Distribution-Free Semi-Supervised Learning with Risk Rewrite ​

Author: Yushi Hirose, Hiroo Irobe, Takafumi Kanamori
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11947v2 Announce Type: new Abstract: Typical semi-supervised learning (SSL) methods rely on distributional assumptions, and their performance degrades when these are violated. While PNU learning, a risk rewriting method, offers a distribution-free alternative, it is restricted to binary c...

📖 Read original article


11. Scale-Aware Attention for Scarce Neural Data: An RG-Flow Transformer on Sleep-EDF EEG ​

Author: Dibakar Sigdel
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11950v1 Announce Type: new Abstract: Brain field potentials are scale-free: their power spectra follow a $1/f^{\beta}$ law whose aperiodic exponent $\beta$ tracks cortical state, and sleep depth in particular is a shift in $\beta$. We ask whether a transformer endowed with an explicit ren...

📖 Read original article


12. Scalable Optimal Transport Algorithm for Network Alignment ​

Author: Elaheh Hassani, Durga Mandarapu, Qi Yu, Hanghang Tong, Ariful Azad
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.DC

arXiv:2607.11952v1 Announce Type: new Abstract: Network alignment identifies node correspondences across different networks and is a fundamental primitive in many data science applications, including social network analysis, fraud detection, and knowledge graph integration. However, state-of-the-art...

📖 Read original article


13. When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary ​

Author: Jim Allchin
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11953v2 Announce Type: new Abstract: Does a reinforcement-learning agent that earns high reward represent its task's latent state, or only a reward-correlated shortcut? Usually this is unanswerable: the "true state" is undefined. We make it exactly answerable with a white-box instrument: ...

📖 Read original article


14. Graph-Constrained Policy Learning for Extreme Clinical Code Prediction ​

Author: Amritpal Singh, Sebastian Torres, Khawar Shakeel, Syed Ahmad Chan Bukhari
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IR

arXiv:2607.11954v1 Announce Type: new Abstract: Clinical code prediction maps unstructured discharge summaries to ICD-10-CM leaf codes in a large, sparse, and deeply hierarchical label space. Most systems treat the task as flat multi-label classification, scoring codes independently and providing li...

📖 Read original article


15. Exact and Certified Data Shapley for Weighted k-Nearest-Neighbor Regression and Soft-Label Prediction ​

Author: Zongye Lyu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DS

arXiv:2607.11956v1 Announce Type: new Abstract: Data Shapley is the standard principled answer to which training points are worth what, and its k-nearest-neighbor (KNN) specialization is the version deployed in practice: the exact estimator shipped by toolkits such as pyDVL and OpenDataVal. Exact al...

📖 Read original article


16. Constructed Reality, Contested Priors: Decoupling and the Architecture of Cognitive Relapse Under the Free Energy Principle ​

Author: MD Ibrahim Hossain Ridoy
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC

arXiv:2607.11958v1 Announce Type: new Abstract: Under the free energy principle, a predictive system does not observe reality directly; it maintains a generative model of the world and experiences that model's best current hypothesis. Can a synthetic environment be made consistent enough that a pred...

📖 Read original article


17. Evaluating Reliability in Machine Learning Models for Early Chronic Kidney Disease Prediction: A Systematic Review of Data Leakage and Predictor Stability ​

Author: Mashrul Hossain, Nafesa Kibria, Fahim Shahriar
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11963v1 Announce Type: new Abstract: The early detection of Chronic Kidney Disease using machine learning has attracted significant interest in healthcare-related computer science. Despite rapid advancements in this field, many reported studies remain inconsistent and potentially misleadi...

📖 Read original article


18. LIDAR-AD: A Decoder-Free Latent-Interaction Dreamer with Action-Residual Chains for Autonomous Driving ​

Author: Yongzhi Liu, Yang Xiao, Zhong Cao, Zeng Kang, Sunan Zhang, Zhaozhi Dong, Guojun Yu, Weichao Zhuang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11964v1 Announce Type: new Abstract: Autonomous driving requires long-horizon closedloop decision making in dynamic traffic environments. Latent world models offer an effective framework for this problem by enabling imagination-based decision making in compact latent spaces. However, mult...

📖 Read original article


19. Beyond Coordinate Gauge: An Audited Protocol for Detecting Donor-Specific Functional Fingerprints after Neural Collapse ​

Author: Truong Xuan Khanh, Phan Thanh Duc
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11967v1 Announce Type: new Abstract: Independently trained neural networks have no shared neuron-index reference frame, so comparing them requires accounting for coordinate freedom. Neural Collapse sharpens this problem: networks converge toward a shared, low-dimensional geometry, raising...

📖 Read original article


20. Self-Evolving In-Context Learning for Direct Pilot-to-Beamformer Design in MU-MISO Systems ​

Author: Yubo Zhang, Xiaodong Wang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, math.IT

arXiv:2607.11970v1 Announce Type: new Abstract: We develop an enhanced in-context learning (ICL) framework to improve the performance of pilot-based beamforming in multi-user multiple-input single-output (MU-MISO) systems. The proposed scheme integrates the ICL-Transformer backbone with the pilot en...

📖 Read original article


21. Learning to Discretize: Diffusion-Based Adaptive Mesh with Spectral Guidance ​

Author: Zixuan Shen (Central South University), Bingchuan Wang (Central South University), Zhi Wang (Nanjing University), Yong Wang (Central South University)
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11974v1 Announce Type: new Abstract: Most neural partial differential equation (PDE) surrogates learn how fields evolve after a grid has already been chosen. However, before any operator is applied, the grid has already determined how modeling capacity is allocated across space, resolutio...

📖 Read original article


22. Signal-Guided Optimization for Machine Unlearning ​

Author: Xujia Li, Dan Li, Jian Lou, Wenjie Feng
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11975v1 Announce Type: new Abstract: Current machine unlearning methods predominantly rely on global, coarse-grained intervention strategies. They lack precise pilot signals to guide the unlearning process and fail to provide differentiable guidance across different unlearning tasks. Due ...

📖 Read original article


23. LiteTopK: Exploiting the Curse of Dimensionality for a Fused Indexer-TopK Kernel in Long-Context Sparse Attention ​

Author: Ziqi Yin, Jianyang Gao, Peiqi Yin, Jiangneng Li, Gao Cong
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.11976v2 Announce Type: new Abstract: Indexer-TopK, the operation to compute the scores and select the top-k candidates, is widely used by sparse attention kernels in large language models and vector retrieval in recommendation systems and vector databases. However, existing GPU-based Inde...

📖 Read original article


24. Gene Expression-Informed Jointly Controlled Generative Modeling for Precision Molecular Design ​

Author: Hang Yuan, Chen Li, Wenjun Ma, Tadahiko Murata, Yuncheng Jiang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11978v1 Announce Type: new Abstract: Precision molecular design aims to discover personalized drug candidates through joint control of multiple conditions, such as biological relevance and molecular design strategies. Biological relevance reflects cellular functional states under disease ...

📖 Read original article


25. Sparse Inter-Layer Dependencies of Transformer FFN Neurons ​

Author: Johannes Knittel, Hanspeter Pfister
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.11990v1 Announce Type: new Abstract: Feedforward network (FFN) blocks account for a large fraction of the parameters and computation in Transformer architectures, yet their internal structure remains difficult to interpret due to the additive superposition induced by the residual stream. ...

📖 Read original article


26. Mitigating The Effect of Class Imbalance in Data with Hierarchical and Dependable Structure ​

Author: Bipin Chhetri, Deepika Giri, Avishek Kadel, Rabin Kumar Karki, Akbar Siami Namin
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11994v1 Announce Type: new Abstract: Classifying cybersecurity vulnerabilities using the Common Weakness Enumeration (CWE) taxonomy is challenging due to extreme class imbalance and strong hierarchical dependencies among weakness categories. Although oversampling techniques such as Synthe...

📖 Read original article


27. Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs ​

Author: Nikita Kozodoi, Zainab Afolabi, Jack Butler
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2607.11997v2 Announce Type: new Abstract: Multi-task model merging combines separately trained expert models into a single model that handles all tasks without co-training. Standard practice merges experts at their optimal validation loss. We challenge this convention by systematically studyin...

📖 Read original article


28. Institutional Equity Holdings Prediction Using Node Affinities of Dynamic Graphs ​

Author: Emad Izadifar, Zahed Rahmati
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12067v1 Announce Type: new Abstract: Institutional equity holdings disclosed in SEC Form 13F filings provide a rich temporal record of portfolio decisions by large investment managers. However, forecasting future allocations and modeling future demand remains challenging due to disclosure...

📖 Read original article


29. Sparse Autoencoders for Interpretable Out-of-Distribution Detection ​

Author: Ayush Karmacharya (Purdue University), Luke Luschwitz (Purdue University), Lucia Romero (Purdue University), Yanan Niu (EPFL), Joseph Campbell (Purdue University)
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.12094v1 Announce Type: new Abstract: Reliable detection of out-of-distribution (OOD) samples is crucial for the safe deployment of machine learning models. Neural networks often produce overconfident predictions for inputs that deviate from their training data, leading to significant degr...

📖 Read original article


30. PFAdapter: Hierarchical LoRA Decomposition for Personalized Federated MLLMs ​

Author: Jing Liu, Kun Yang, Yan Wang, Dingkang Yang, Xiaoshuai Hao, Wei Zhang, Yang Liu, Wei Zhou
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC, cs.MA, cs.NI

arXiv:2607.12111v1 Announce Type: new Abstract: Agentic AI systems are reshaping communications and networking by deploying autonomous intelligent agents capable of collaborative learning while maintaining data privacy at network edges. Within distributed network environments, Multimodal Large Langu...

📖 Read original article


31. Continual Learning with Elastic Regularization and Synthetic Replay for Federated MLLM Fine-Tuning ​

Author: Jing Liu, Chenxuanyin Zou, Jiayang Ren, Gaoyun Fang, Chengfang Li, Yan Wang, Zhenchao Ma, Bo Hu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, cs.DC

arXiv:2607.12112v1 Announce Type: new Abstract: Federated fine-tuning of Multimodal Large Language Models (MLLMs) across distributed networks enables privacy-sensitive adaptation to evolving data streams, yet a fundamental obstacle prevents robust deployment in dynamic environments: catastrophic for...

📖 Read original article


32. An Agentic AI Scientific Community for Automated Neural Operator Discovery ​

Author: Luis Loo, Ulisses Braga-Neto
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12122v1 Announce Type: new Abstract: We present an agentic approach to autonomous neural operator discovery based on an AI scientific community, which consists of a swarm of virtual laboratories that interact under a citation-based economy of influence. Highly-cited labs found new labs th...

📖 Read original article


33. From Geometric Recovery to Causal Validation: A Reproducible Audit of Sparse Autoencoder Features, from Superposition Geometry to Causal Inertness ​

Author: Mohamed Abdessalem Bal
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12166v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) are the standard for decomposing superposed neural representations into interpretable features, and evaluation relies predominantly on correlational recovery metrics -- cosine similarity between ground-truth directions and de...

📖 Read original article


34. Forgetful Attention: A Trainable Support-Vector Memory with Certified Selection and Exact Unlearning ​

Author: Vishwajith Ramesh
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12204v1 Announce Type: new Abstract: Attention can be viewed as an online learner over context, yet existing test-time memories cannot certify that dropping a token leaves outputs unchanged or delete its influence outright. We introduce Support Vector Attention (SV-Attention), a max-margi...

📖 Read original article


35. Speculate with Memory: Lossless Acceleration for LLM Agents ​

Author: Yu Li, Qinyuan Ye, Prafulla Kumar Choubey, Jiaxin Zhang, Chien-Sheng Wu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.12236v1 Announce Type: new Abstract: Speculative execution accelerates LLM agents by using a smaller, cheaper model to predict and pre-launch the next step while the environment is idle. However, existing speculators are stateless and discard all information between tasks, preventing pred...

📖 Read original article


36. Cluster-Weighted EDMD ​

Author: Lorenzo Tomaz, Judd Rosenblatt, Flavio Kicis, Thomas B. Jones, Diogo Schwerz de Lucena
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.12243v1 Announce Type: new Abstract: Extended Dynamic Mode Decomposition (EDMD) approximates Koopman operators from data, but a single global operator is inefficient when different state-space regions exhibit distinct local dynamics. We introduce Cluster-Weighted EDMD (CW-EDMD), which joi...

📖 Read original article


37. Proximity Features: Privacy-Compliant Cold-Start Personalization at Airbnb ​

Author: Wei Jiang, Bin Xu, Hui Gao, Bharathi Thangamani, Weiwei Guo, Sundar Srinivasavaradhan, Tracy Yu, Huiji Gao, Michael Kinoti
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12246v1 Announce Type: new Abstract: Personalization in two-sided marketplaces relies heavily on user-level features, yet for platforms with infrequent, high-consideration purchases, a large fraction of users lack sufficient history for effective recommendation, spanning both paid and org...

📖 Read original article


38. Understanding Structured Health Data through Interaction-Aware Mixture-of-Experts ​

Author: Ji Hwan Park, Ying Ding, Tianjin Guo
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12255v1 Announce Type: new Abstract: We study interaction-aware mixture-of-experts for post-stroke rigidity prediction using multi-level views of structured health records. Despite minimal performance gains, routing attribution reveals systematic importance differences across views, under...

📖 Read original article


39. From Many to Meaningful: Feature-Guided Zero-Shot Chronic Kidney Disease Screening Using Large Language Models ​

Author: Muhammad Ashad Kabir, Sirajam Munira
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12260v1 Announce Type: new Abstract: Early screening of chronic kidney disease (CKD) is essential for preventing irreversible progression; however, many machine learning (ML)-based screening methods remain difficult to deploy in community and resource-limited screening settings due to the...

📖 Read original article


40. Saturation Makes Quantization Error Additive: A Coverage Model with a Certificate ​

Author: Joshua Hill
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12266v1 Announce Type: new Abstract: Mixed-precision quantization must decide which parts of a model to keep at higher precision. A common premise, shared by sensitivity-based methods such as HAWQ and CoopQ, is that the loss from quantizing a set of layers can be reconstructed from per-la...

📖 Read original article


41. Track, Rank, Crack: Epistemic Working Memory Scales Multi-Hop Reasoning in Language Agents ​

Author: Ning Liu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.12267v1 Announce Type: new Abstract: Language agents that interleave reasoning and tool use degrade sharply as reasoning chains lengthen, even when each individual step is easy. We trace this to context dilution: an agent's investigative state (what it has confirmed, what it suspects, and...

📖 Read original article


42. Quantum Port-Hamiltonian Neural Networks: Learning Conservative and Dissipative Dynamics via Measurement-Induced Nonlinearity ​

Author: Dibakar Sigdel
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12269v1 Announce Type: new Abstract: We introduce Quantum Port-Hamiltonian Neural Networks (Q-pHNNs), a family of parameterised quantum circuits that learn classical dynamics in a structure-preserving manner. The framework relies on the Isomorphic Hamiltonian Mapping (IHM): the skew-symme...

📖 Read original article


43. A hybrid analytical-PINN model for subsurface simulation of geothermal heat exchangers in heterogeneous underground ​

Author: Moke Rao, Thomas Hamacher, Smajil Halilovic
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12271v1 Announce Type: new Abstract: In this paper, a parametric physics-informed neural network for solving the heterogeneous soil thermal problem with borehole heat exchangers (BHEs) as singular sources is developed. There are three novel features in the present framework; namely, (i) t...

📖 Read original article


44. Gradient Flow Dynamics and Implicit Bias of Diagonal Linear Networks under Infinitesimal Initialization ​

Author: Jiajie Zhao, Jianxing Wang, Junjie Yang, Zhiwei Bai, Yaoyu Zhang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12332v1 Announce Type: new Abstract: We study the gradient flow dynamics of diagonal linear networks for regression tasks under infinitesimal initialization. Extending Theorem 1 from Pesme & Flammarion (2023), we generalize the analysis to both deep diagonal linear networks and a broader ...

📖 Read original article


45. Generating Developable 3D Molecules via Pocket-Conditioned Diffusion and Property-Aware Optimization ​

Author: Ruoxi Gao, Jiangweizhi Peng, Ziqi Chen, Frazier N. Baker, David C. Kombo, John L. Kane Jr., Andrew A. Scholte, Yi Li, Matthew J. LaMarche, Luigi I. Iconaru, Hans-Peter Biemann, Mingyi Hong, Xia Ning
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12349v1 Announce Type: new Abstract: Drug discovery and development is time-consuming and resource-intensive, motivating computational approaches such as diffusion models for de novo drug design. Many such models follow the structure-based drug design (SBDD) paradigm, generating molecules...

📖 Read original article


46. Reducing information dependency does not cause training data privacy. Adversarially non-robust features do ​

Author: Rasmus Torp, Shailen K. Smith, Adam Breuer
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12354v1 Announce Type: new Abstract: In this paper, we challenge the prevailing view that information dependency (including rote memorization) drives training data exposure to image reconstruction attacks. We show that extensive exposure can persist without rote memorization and is instea...

📖 Read original article


47. Same Loss, Same Noise, Opposite Schedules: Noise Structure and Optimizer Normalization Jointly Determine Whether Learning-Rate Cooldown Helps ​

Author: Subham Singh, Ashutosh Mishra, Subha Raut
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12360v1 Announce Type: new Abstract: The cooldown phase of a warmup-stable-decay (WSD) learning-rate schedule, now a default in large-model pretraining, lowers the final training loss in some settings and does nothing in others. We give a provable account of which case obtains, and it tur...

📖 Read original article


48. SinAE: A Single-Architecture Flow-Matching Autoencoder for Cross-Domain Atomic Systems ​

Author: Yuxuan Ren, Fan Yang, Jianhua Yao, Yatao Bian
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12380v1 Announce Type: new Abstract: Small molecules, crystals, and proteins all reduce to atoms in 3D space, yet their generative pipelines remain fragmented across domains, each with its Small molecules, crystals, and proteins all reduce to atoms in 3D space, yet their generative pipeli...

📖 Read original article


49. Differentiable Clone-Structured Causal Graphs for End-to-End Cognitive Map Learning from Image Sequences ​

Author: Arash Nikzad, Sasan Sarbishegi, Ali Dasmeh, Muhammad Asif, Parsa Gharavi, Erik Husom, Sagar Sen, Andrew B. Lehr, Olivier Penacchio, Ana Clemente, Tristan M. St"ober
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, q-bio.NC

arXiv:2607.12382v1 Announce Type: new Abstract: How can an agent build a structured map of its world from nothing but an ongoing sequence of raw sensory input and its own movements, especially when natural variation means exact sensory patterns rarely repeat? The Clone-Structured Causal Graph algori...

📖 Read original article


50. ReDiTT: Retrieval Augmented Conditional Diffusion Transformers for Asynchronous Time Series ​

Author: Saiyue Lyu, Zhitian Zhang, Ruizhi Deng, Thibaut Durand
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12391v1 Announce Type: new Abstract: We present a diffusion based model for asynchronous time series prediction, where the goal is to predict the next inter event time and event type. To address the inherent uncertainty of future events, we introduce ReDiTT, a retrieval augmented conditio...

📖 Read original article


51. Mechanical Analysis of Parachute Suspension Line Deployment with Binding Tapes Using PINN ​

Author: Xiang Zhao, Ronghui Quan, Yaqi Xiao, Junlin Chen
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12409v1 Announce Type: new Abstract: Parachutes are widely utilized in aviation, aerospace and lifesaving missions. As the initial stage of parachute deployment, suspension line extraction and straightening directly determines the smooth implementation of subsequent inflation procedures. ...

📖 Read original article


52. PolarBM: Complex-valued Boltzmann Machine for Modeling Audio Signals in Polar and Log-polar Coordinates ​

Author: Toru Nakashika, Kohei Yatabe
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.SD, eess.AS, stat.ML

arXiv:2607.12417v1 Announce Type: new Abstract: Although vast amounts of data, such as audio signal spectra, are naturally represented using complex numbers, conventional machine learning methods often simplify complex-domain problems by employing frameworks designed for real-valued variables. While...

📖 Read original article


53. Fisher Rank Inflation: A Spectral Signature of Memorization under Label Noise ​

Author: Satwik Bathula, Anand A. Joshi
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.12438v1 Announce Type: new Abstract: Deep networks trained with label noise often learn clean structure before memorizing corrupted labels. We show that this transition leaves a spectral signature in the centered scatter of per-example last-layer gradients. Its effective rank transiently ...

📖 Read original article


54. The Computational Basis of Confidence in Large Language Models ​

Author: Dharshan Kumaran, Viorica Patraucean, Maks Ovsanikov, Petar Veli\v{c}kovi'c, Nathaniel Daw
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.12447v1 Announce Type: new Abstract: Reliable confidence -- the probability that a model's own answer is correct -- is essential for the trustworthy deployment of language models. Existing work has largely evaluated confidence by how well it predicts correctness and whether it is calibrat...

📖 Read original article


55. Exploring Zero-Shot Foundation Models for Multivariate Time Series Anomaly Detection ​

Author: Martin Uray, Saverio Messineo, Roland Kwitt, Stefan Huber
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12454v1 Announce Type: new Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is essential for reliability and safety in domains such as industrial process monitoring and financial risk management, yet conventional approaches rely on application-specific models that are costly t...

📖 Read original article


56. Sample Efficient Generative Optimization for Molecular Design ​

Author: Sarina Kopf, Cristina Nevado, Philippe Schwaller
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12488v1 Announce Type: new Abstract: Molecular optimization in drug discovery, materials design, and catalysis requires searching vast chemical spaces under tight evaluation budgets, since high-fidelity oracles and experimental measurements are costly. The practical impact of an optimizat...

📖 Read original article


57. Adversarial Attacks on Online Handwriting using Salience-based Temporal Editing ​

Author: Yataro Tamura, Brian Kenji Iwana, Jiseok Lee
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.12500v1 Announce Type: new Abstract: Deep learning models for online handwriting recognition have been shown effective and are increasingly deployed in practical applications. However, their vulnerability to adversarial attacks is still a challenge. Existing adversarial methods are predom...

📖 Read original article


58. What Does Goodness Measure? A Likelihood-Ratio Account of Forward-Forward Learning ​

Author: Paolo Giannitrapani
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, eess.IV, stat.ML

arXiv:2607.12501v2 Announce Type: new Abstract: The Forward-Forward (FF) algorithm trains each layer locally, so that a scalar goodness - the sum of squared activations - is high on real inputs and low on contrastive ones, with activations normalized between layers. Both choices are usually treated ...

📖 Read original article


59. OOD-RL-Bench: A Benchmark Framework for Out-of-Distribution Detection in Reinforcement Learning ​

Author: Emil Mittag, Richard Dazeley, Peter Vamplew
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.12523v1 Announce Type: new Abstract: Reliable reinforcement learning (RL) agents must maintain operational integrity amidst sensor malfunctions, dynamic disturbances, and slow environmental shifts. The detection of out-of-distribution conditions is pivotal to determining when an agent's o...

📖 Read original article


60. From Preimage Search To Source-Grounded Feature Inversion ​

Author: Kaixiang Shu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12526v1 Announce Type: new Abstract: Interpreting a neural network requires understanding what its internal features extract from a particular input. Feature inversion seeks to express a selected feature in the input domain, but canonical iterative methods search for an input whose re-enc...

📖 Read original article


61. A JoLT for the KV Cache: Near-Lossless KV Cache Compression via Joint Tucker and JL-Residual Allocation for LLMs ​

Author: Rahul Krishnan, Volker Schulz
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, math.OC

arXiv:2607.12550v1 Announce Type: new Abstract: The key-value (KV) cache has become the dominant memory cost of transformer inference. It grows with batch size, context length, and depth, and at long context it, rather than the model weights, sets the ceiling on throughput. Two families of methods r...

📖 Read original article


62. Lightweight Multi-Scale Anomaly Detection for Resource-Constrained Edge Devices ​

Author: Raheen Junaid Wani, Smruti R. Sarangi
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12599v1 Announce Type: new Abstract: Time-series anomaly detection is increasingly important in IoT systems, sensor networks, and edge monitoring applications, where models must operate under strict constraints on memory, latency, and power consumption. While recent deep-learning approach...

📖 Read original article


63. The Geometry of Memorization: Finite-Time Spectral Sensitivity as a Diagnostic for Flow Matching Models ​

Author: Shuchan Wang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12616v1 Announce Type: new Abstract: Continuous-time generative frameworks construct probability paths between base and target domains by optimizing time-dependent velocity fields. While theoretical targets favor straight trajectories, empirical networks develop complex path deformations....

📖 Read original article


64. Learning Forced Multibody Dynamics on Lie Groups ​

Author: Martine Dyring Hansen, Marta Ghirardelli, Elena Celledoni, David Martin de Diego, Brynjulf Owren
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, math.DS, math.SG

arXiv:2607.12627v1 Announce Type: new Abstract: We propose an architecture for learning the dynamics of mechanical systems based on discrete forced Euler-Lagrange equations on Lie groups using only position data. By formulating the dynamics directly on manifold-valued configuration spaces, the metho...

📖 Read original article


65. AdaPCLA: Adaptive Prior-Calibrated Logit Adjustment for Long-Tailed Longitudinal EHR Generation ​

Author: Shuai Cui, Chen Wenxuan, Wenjie Du, Jian Lou, Dan Li, Wenjie Feng
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12645v1 Announce Type: new Abstract: Generative modeling of longitudinal Electronic Health Records is increasingly important for privacy-preserving research, yet standard autoregressive models tend to underrepresent the co-occurrence structure of tail events (i.e., diseases, symptoms), re...

📖 Read original article


66. Extractable Memorization From First Principles ​

Author: A. Feder Cooper, Marika Swanberg, Jamie Hayes, Lea Duesterwald, Christopher De Sa, Daniel E. Ho, Mark A. Lemley, Percy Liang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.12649v1 Announce Type: new Abstract: Recent work on extractable memorization in LLMs suffers from two contrasting validity problems. Some studies overstate extraction, e.g., relying on sequences too short to distinguish memorization from predictability. Others imply that extraction is unr...

📖 Read original article


67. Evidence-Grounded Verified Agentic Reasoning: A Path Toward Eliminating LLM Hallucination in Empirical Inference via Tool-Attested Kernel Proofs ​

Author: Junyu Ren
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CY, cs.SE

arXiv:2607.12650v1 Announce Type: new Abstract: Tool access alone does not make LLM empirical reasoning governable: accepted outputs need not descend from attested evidence, and accepted deductions need not hold up under formal scrutiny. We present EG-VAR (Evidence-Grounded Verified Agentic Reasonin...

📖 Read original article


68. Learning-based Probabilistic Load Forecasting with Post-hoc and In-model Uncertainty ​

Author: Sarah Al-Shareeda, Gulcihan Ozdemir, Heung Seok Jeon
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.12730v1 Announce Type: new Abstract: Smart-building load forecasters are often trained offline on dense, multivariate, high-frequency data, but deployment may provide only hourly, feature-limited inputs. Missing features must then be reconstructed, and their errors can propagate through t...

📖 Read original article


69. What Makes a Representational Prior Work? Feature Families, Label-Free Invariances, and Critical Windows in Grokking ​

Author: Gunner Levi Howe
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12735v1 Announce Type: new Abstract: Companion work showed the grokking delay is causally the time to form task-structured representations, injectable via a contrastive prior. Here we characterize what makes such a prior work, across four axes, in 188 new runs. Content: a coherent, learna...

📖 Read original article


70. Constraint-Aware Aggregation for Federated Reinforcement Learning in Microgrid Energy Coordination ​

Author: Usman Haider, Karl Mason
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.12763v1 Announce Type: new Abstract: Federated Reinforcement Learning (FedRL) enables coordination of distributed energy resources without sharing raw local data, but standard aggregation methods such as FedAvg do not account for system-level constraints, often leading to unsafe global be...

📖 Read original article


71. Learning Mechanistic Reasoning for Chemical Reactions with Large Language Models ​

Author: Xingyu Dang, Haocheng Tang, Junmei Wang, Yanjun Li
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CE, cs.CL, q-bio.BM

arXiv:2607.12771v1 Announce Type: new Abstract: Reaction mechanisms consist of the step-by-step sequences of elementary reactions that explain chemical transformations. Learning the mechanism logic is therefore essential for enhancing the fundamental chemical intelligence of large language models (L...

📖 Read original article


72. AVQ-Attention: Adaptive Vector-Quantized Attention ​

Author: Winfried van den dool, Patrick Forr'e, Amir Habibian, Yuki M. Asano, Max Welling
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.12789v1 Announce Type: new Abstract: The $\mathcal{O}(N^2)$ complexity of attention over $N$ tokens remains a computational bottleneck in transformer models. Vector-Quantized (VQ) attention reduces this to $\mathcal{O}(MN)$ by representing keys with $M$ codewords, but applies uniform code...

📖 Read original article


73. Accelerating Masked Diffusion Large Language Models: A Survey of Efficient Inference Techniques ​

Author: Daehoon Gwak, Minhyung Lee, Junwoo Park, Jaegul Choo
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.12829v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer a theoretical advantage in parallel generation over standard autoregressive models. However, parallel generation alone does not guarantee practical speedups. Realizing this efficiency requires specialized i...

📖 Read original article


74. Verifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage Control ​

Author: Takumi Shioda, Kohei Terashima, Tatsuo Nagai
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12856v1 Announce Type: new Abstract: Buildings are expected to shift cooling loads in response to grid conditions. Thermal energy storage (TES) enables this shift, but scheduling it well requires planning hours ahead under storage constraints. Model predictive control (MPC) and reinforcem...

📖 Read original article


75. Energy-Based Physics-Informed Form Finding for Clustered Tensegrity Structures ​

Author: Jing Qin, Muhao Chen
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2607.12888v1 Announce Type: new Abstract: Tensegrity form-finding and physical property prediction are fundamental inverse problems in structural mechanics, which aim to determine equilibrium configurations and internal force distributions. These problems are challenging due to strong nonlinea...

📖 Read original article


76. Contrastive-Collapsed Loss for Flexible and Geometrically Optimal Embeddings and Faster Convergence ​

Author: Blanca Cano-Camarero, 'Angela Fern'andez-Pascual, Jos'e R. Dorronsoro
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12916v1 Announce Type: new Abstract: In this work, we introduce CoCo, a loss function aimed at learning normalized and well-structured representations. The proposed loss encourages intra-class collapse and inter-class contrast while preserving sufficient flexibility for neural networks to...

📖 Read original article


77. Efficient Sequential Calibration with $O(T^{2/3-\epsilon})$ Error Bound ​

Author: Zihan Zhang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.12928v1 Announce Type: new Abstract: We study the online binary sequential calibration problem. A recent breakthrough by \citet{dagan2024breaking} overcomes the classical (T^{2/3}) barrier for calibration error. Building on this result, we present an efficient randomized forecaster that...

📖 Read original article


78. The Spectrum Is Not Enough: When Context Helps Time-Series Forecasting ​

Author: Mert Onur Cakiroglu, Mehmet Dalkilic, Hasan Kurban
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.13006v2 Announce Type: new Abstract: A growing family of indices scores how predictable a series is from its spectrum. Practitioners increasingly read these scores as answering a different question: whether \emph{adding context}, a longer lookback, a retrieval plug-in, or a pretrained mod...

📖 Read original article


79. TerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale ​

Author: Zhouchonghao Wu, Akshay Rangesh, Weixin Li, Wei-Jer Chang, Zachary Lee, Tim Wang, Wei Zhan
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.RO

arXiv:2607.13028v1 Announce Type: new Abstract: Training robust autonomous driving agents requires a simulator that is fast enough for reinforcement learning at scale, realistic enough to ground behavior in real-world map structure, and diverse enough to cover the safety-critical long tail that logg...

📖 Read original article


80. The Seriality Gap in Video Diffusion Models ​

Author: Jorge Diaz Chao, Konpat Preechakul, Yuxi Liu, Yutong Bai
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2607.13031v1 Announce Type: new Abstract: When one ball strikes another, then another, video models should predict the consequences of each bounce. In controlled experiments on multi-ball hard-sphere dynamics, we find that the performance of standard bidirectional video diffusion degrades as t...

📖 Read original article


81. TAKE: Trajectory-Aware Knowledge Estimation for Text Dataset Distillation ​

Author: Tri-Nhan Vo, Dang Nguyen, Sunil Gupta
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.11898v1 Announce Type: cross Abstract: Large-scale text corpora have become a quiet bottleneck in modern NLP, not just in storage, but in the accumulated cost of training, fine-tuning, and continual learning. We propose a text dataset distillation framework that reduces corpora to as litt...

📖 Read original article


82. Predictive Modeling of High-Altitude Clear Air Turbulence in the United States: A Machine Learning Approach ​

Author: Kadir Gokdeniz, Irem Ulku
Published: 7/15/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG

arXiv:2607.11899v1 Announce Type: cross Abstract: High-altitude Clear Air Turbulence (CAT) poses significant risks to aviation safety due to its unpredictability and challenges in detection. This study leverages machine learning models to improve CAT prediction within U.S. airspace at 200-350 hPa pr...

📖 Read original article


83. Benchmarking Sensor Robustness in Plasma Diagnostic Models: A Systematic Evaluation on TokaMark ​

Author: Neerav Gupta
Published: 7/15/2026, 4:00:00 AM
Categories: physics.plasm-ph, cs.LG

arXiv:2607.11915v1 Announce Type: cross Abstract: Plasma diagnostic models for tokamak fusion devices are almost universally evaluated on clean, complete sensor data. In practice, fusion diagnostics fail regularly: acquisition systems start late, individual sensors die, and signal dropouts cluster p...

📖 Read original article


84. AAAI-26 Dual Submissions: Novel Challenges ​

Author: Kiri L. Wagstaff, Joydeep Biswas, Erich Merrill III, Bo An, Ida Camacho, David J. Crandall, Matthew E. Taylor
Published: 7/15/2026, 4:00:00 AM
Categories: cs.DL, cs.AI, cs.CY, cs.LG

arXiv:2607.11918v1 Announce Type: cross Abstract: Dual submissions, in which identical or substantially similar papers are simultaneously submitted to one or more archival venues, without cross-citation or disclosure, are a growing problem for the AAAI Conference and other scientific publication ven...

📖 Read original article


85. Do You Remember? Toward Memory-Centric Multimodal AI ​

Author: Xuguang Yu, Weigang Zheng, Minyue Yu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.CV, cs.LG

arXiv:2607.11919v1 Announce Type: cross Abstract: Human memory is reconstructive, not a faithful recording. Current multimodal LLMs (MLLMs) lack this capability: they process images through a frozen visual encoder, produce a one-shot text output, and discard internal representations. We present DoYo...

📖 Read original article


86. Near-Optimal Learning of Gaussian Sobolev Operators ​

Author: Ben Adcock, Michael Griebel, Gregor Maier
Published: 7/15/2026, 4:00:00 AM
Categories: math.NA, cs.IT, cs.LG, cs.NA, math.IT

arXiv:2607.11921v1 Announce Type: cross Abstract: A key question in operator learning is how to design surrogate operators with provable approximation guarantees in reasonable computational time. Whereas smooth operators can be approximated efficiently, i.e., with at least algebraic convergence in t...

📖 Read original article


87. Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking ​

Author: Shreeya Dasa Lakshminath, Shubhan S
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.IR, cs.LG

arXiv:2607.11933v1 Announce Type: cross Abstract: Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-time deployment. We address this by fine-tuning LLaMA 3 (8B) as a drop-in reranker using a two-stage...

📖 Read original article


88. GenDiff: A Dose and Anatomy Aware Diffusion Model with Structural Prior Refinement for Low-Dose CT Reconstruction and Generalization ​

Author: Md Imam Ahasan, Guangchao Yang, A F M Abdun Noor, Kah Ong Michael Goh, S. M. Hasan Mahmud, Md Mahfuzur Rahman
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.11941v1 Announce Type: cross Abstract: Computed tomography (CT) is a critical imaging modality for clinical diagnosis, but reducing radiation dose inevitably introduces severe noise and structured artifacts that degrade image quality. Existing deep learning-based low-dose CT (LDCT) recons...

📖 Read original article


89. MAGE: Understanding Stability-Performance Trade-offs in Multi-component Prompt Optimization ​

Author: Prateek Singh
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.11944v1 Announce Type: cross Abstract: How do different components of iterative prompt optimization interact, and what happens when they are combined? We investigate this through MAGE (Memory-Augmented Goal-directed Prompt Evolution), a controlled analysis framework for studying component...

📖 Read original article


90. Belief-reality separation lives in routing over a shared value slot in language models ​

Author: Oliver Steele, Jiangtao Wen, Yuxing Han
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.11945v1 Announce Type: cross Abstract: Capable language models hold what a character believes apart from what is true: told "Anna believes the cup is blue; in reality it is red," they answer blue about Anna and red about the world. Where in the computation does that separation live? We sh...

📖 Read original article


91. Ontology-Amplified Distillation and Contextuality Auditing for Sovereign Enterprise Language Models: A Combined Proof-of-Mechanism and Negative-Results Method Study ​

Author: Thanh Luong Tuan
Published: 7/15/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2607.11948v1 Announce Type: cross Abstract: Regulated financial institutions operating under data-residency rules need tenant-owned language models that can run inside the institution's perimeter. This paper combines two related FAOS studies into one mechanism-and-control article. First, it re...

📖 Read original article


92. Contrastive Joint-Embedding Prediction for Representation Learning in Structural MRI ​

Author: Fabian Mager, Lars Kai Hansen
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.11962v1 Announce Type: cross Abstract: Self-supervised learning offers a compelling approach for medical imaging, where labeled data are scarce and acquisition costs are high. We present COJEPA, a self-supervised framework for volumetric brain MRI that combines a joint-embedding predictiv...

📖 Read original article


93. Did We Actually Fix It? An Independent Adversarial Stress-Test of Post-Point-Adjustment Evaluation Metrics for Time-Series Anomaly Detection ​

Author: Zongye Lyu
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.11969v1 Announce Type: cross Abstract: Point-adjustment (PA), long the default scoring protocol in time-series anomaly detection (TSAD), was shown by Kim et al. (2022) to award near-perfect F1 to random scores. The field migrated to replacement metrics: PA%K, range-based precision/recall,...

📖 Read original article


94. Removable Defects: The Economics and Limits of Deliberate Deficiency ​

Author: Cheng Qian
Published: 7/15/2026, 4:00:00 AM
Categories: econ.EM, cs.AI, cs.LG, stat.ML

arXiv:2607.11983v2 Announce Type: cross Abstract: A specialist tolerates blind spots that a generalist does not. Usually this is treated as a cost to be minimized. We treat it as a design variable: a deficiency can be kept because it pays and removed on demand in the rare situation where it would be...

📖 Read original article


95. VQCSim: When Does Compile-Once Statevector Simulation Beat Generic Quantum Frameworks? ​

Author: Anton Firc, Martin Pere\v{s}'ini, Vojt\v{e}ch Mr'azek, Kamil Malinka, Vojt\v{e}ch Stan\v{e}k, Zbyn\v{e}k Li\v{c}ka, Nouhaila Innan, Walid El Maouaki, Alberto Marchisio, Muhammad Shafique
Published: 7/15/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG

arXiv:2607.11985v2 Announce Type: cross Abstract: Hybrid quantum-classical machine learning workflows repeatedly evaluate many small parametrized circuits during training and model exploration. In this regime, framework dispatch and orchestration overhead often dominate runtime. Prior simulators acc...

📖 Read original article


96. SpikeDS: Dual Sparsity Spikformer for Perineural Invasion Prediction in 3D MRI ​

Author: Induk Um, Youngung Han, Kyeonghun Kim, Yului Jeong, Jina Jeong, Hyunsu Go, Dohyun Kweon, Sungha Park, Junga Kim, Anna Jung, Suah Park, Hyuk-Jae Lee, Pa Hong, Woo Kyoung Jeong, Won Jae Lee, Ken Ying-Kai Liao, Nam-Joon Kim
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.11986v1 Announce Type: cross Abstract: Perineural invasion (PNI) is associated with poor prognosis in cholangiocarcinoma (CCA). However, its detection from 3D MRI remains challenging due to the subtle and spatially heterogeneous imaging signatures at the tumor periphery. Capturing such sp...

📖 Read original article


97. HPC-Enabled Video-based Coastal Wave Parameter Estimation Using V-JEPA and Deep Spatiotemporal Learning ​

Author: Abubakar Hamisu Kamagata, Dharm Singh Jat, Attlee Munyaradzi Gamundani, Saravanakumar Paramasivam, Babangida Sani, Aliyu Zakariyya
Published: 7/15/2026, 4:00:00 AM
Categories: eess.IV, cs.AI, cs.LG

arXiv:2607.11998v1 Announce Type: cross Abstract: High deployment cost, poor spatial coverage and susceptibility to storm conditions are all challenges faced by traditional in-situ methods. This paper presents a video-based and high performance computing (HPC) enabled deep learning framework for joi...

📖 Read original article


98. Learning the Graphical Nature of Symmetries ​

Author: Rashid Barket, Enrico Grimaldi, Yacoub Hendi, Edward Hirst, Adam Onus, Harmeet Singh
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.CO, math.GR

arXiv:2607.12026v1 Announce Type: cross Abstract: Finite groups are rigid algebraic objects, whose Cayley graphs expose a rich network geometry through which group-theoretic structure can be measured, compared, and learned. In this paper, a dataset of $131{,}406$ Cayley graphs is constructed, coveri...

📖 Read original article


99. SymbOmni: Evolving Agentic Omni Models via Symbolic Concept Learning ​

Author: Jinxiu Liu, Jianru Li, Tanqing Kuang, Xuanming Liu, Kangfu Mei, Yandong Wen, Weiyang Liu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.12042v1 Announce Type: cross Abstract: Visual generation is increasingly ubiquitous in diverse domains, from text-to-image/video synthesis to multimodal interactive creation. Yet prevailing monolithic models remain fundamentally constrained by their inability to learn cumulatively and evo...

📖 Read original article


100. Analyzing Image Encoder Choices and Graph Homophily in GCN Frameworks for Breast Ultrasound Classification ​

Author: Sabahattin Mert Daloglu, Ceren Coskun, Harvey Castro, Soner Hacihaliloglu, Ilker Hacihaliloglu
Published: 7/15/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2607.12054v2 Announce Type: cross Abstract: Breast ultrasound is widely used for screening, yet automated analysis remains challenging due to speckle noise, acquisition variability, and weak separation of benign and malignant cases in standard ultrasound imaging. Graph convolutional networks (...

📖 Read original article


101. Learning from Complementary Ultrasound Representations for Liver Disease Classification ​

Author: Sabahattin Mert Daloglu, Gokce Bekar, Ceren Coskun, Senanur Sahin, Harvey Castro, Soner Hacihaliloglu, Halley P. Letter, Ilker Hacihaliloglu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.12062v2 Announce Type: cross Abstract: Differentiating non-alcoholic steatohepatitis (NASH) from non-alcoholic fatty liver disease (NAFLD) using ultrasound remains challenging due to subtle tissue alterations and the limited information available in conventional B-mode imaging. In this wo...

📖 Read original article


102. Dynamic Online Processor-Native Inference for State Estimation ​

Author: Orestis Kaparounakis
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, eess.SP

arXiv:2607.12095v1 Announce Type: cross Abstract: Sensor-rich data-driven applications increasingly use Bayesian approaches to infer latent states of dynamic systems from noisy sensor measurements and physical models. Yet the computation of the likelihood remains an essential bottleneck for accurate...

📖 Read original article


103. TraceSynth: Generating Production-Quality Kernel Traces with Constraint-Guided Diffusion Models ​

Author: Yuvraj Sehgal, Sneh Patel, Mahsa Panahandeh, Naser Ezzati-Jivan, Francois Tetreault
Published: 7/15/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2607.12104v1 Announce Type: cross Abstract: Machine learning models for system diagnostics rely on kernel execution traces to capture fine-grained system behavior, but collecting production traces in industrial systems is costly due to runtime overhead, storage demands, and privacy constraints...

📖 Read original article


104. Robust In-Hand Manipulation via Priors in Reinforcement Learning and Mechanical Design ​

Author: Yifei Chen, Shihan Lu, Ed Colgate, Kevin Lynch
Published: 7/15/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.12105v1 Announce Type: cross Abstract: In-hand manipulation without external sensing is challenging due to uncertainties from finger-object contacts and disturbances by gravity. While reinforcement learning has shown promise in learning complex finger gaiting, existing approaches do not p...

📖 Read original article


105. FlashDiff: Efficient Regional Execution and Scheduling for Diffusion Model Serving ​

Author: Yaqi Qiao, Ping He, Songrun Xie, Ayush Barik, Chensong Zhang, Zhengzhong Tu, Fan Lai
Published: 7/15/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2607.12121v1 Announce Type: cross Abstract: Diffusion models have become the central backbone for modern image, video, and audio generation, but their efficient service remains a challenge. Unlike autoregressive decoding, diffusion inference repeatedly updates high-dimensional spatial or tempo...

📖 Read original article


106. Generating Physically Plausible Parachute Dynamics with Deep Generative Modeling ​

Author: Yulong Yang, Clara O'Farrell, Christine Allen-Blanchette
Published: 7/15/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2607.12143v1 Announce Type: cross Abstract: Accurately modeling the dynamics of planetary parachute and entry vehicle systems is critical for Entry, Descent, and Landing events such as vehicle separation and sensor activation. These dynamics are difficult to capture with traditional system-ide...

📖 Read original article


107. Falsifying Causal Graphs With Outlier Events ​

Author: William Roy Orchard, Philipp M. Faller, Dominik Janzing
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.12145v1 Announce Type: cross Abstract: True causal relationships are rarely known, and inferring causal graphs from data is hard. A fundamental challenge is how to assess whether a given causal graph is good in the absence of a ground truth. We propose falsifying candidate causal graphs b...

📖 Read original article


108. Self-Consistent Flow: Unifying Velocity and Endpoint Prediction for Rectified Flow Models ​

Author: Xu Han, Jiajing Hu, Li-Ping Liu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2607.12171v1 Announce Type: cross Abstract: In rectified-flow-based generative models, the neural network can be trained to predict two different targets, such as the instantaneous velocity or the data endpoint, to perform denoising. Although prior work shows that these parameterizations lead ...

📖 Read original article


109. Decentralized Gradient Descent: Bottleneck Regimes and Budget Complexity ​

Author: Nicol`o Michelusi
Published: 7/15/2026, 4:00:00 AM
Categories: cs.DC, cs.LG, eess.SP

arXiv:2607.12172v1 Announce Type: cross Abstract: Decentralized gradient descent (DGD) is widely used for solving distributed optimization problems over networks of agents. While its convergence properties are well understood, less is known about the communication and computation resources required ...

📖 Read original article


110. Fine-Tuned Multi-Agent Framework for Detecting OCEAN in Life Narratives ​

Author: Rasiq Hussain, Darshil Italiya, Joshua Oltmanns, Mehak Gupta
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.MA

arXiv:2607.12215v1 Announce Type: cross Abstract: Accurately assessing personality from text is challenging because traits are latent, context-dependent, and often subtly expressed across long narratives. Large language models (LLMs) offer new opportunities by processing extensive textual contexts, ...

📖 Read original article


111. Gradient-Free Topology Adaptation for Power Flow Surrogates via In-Context Whitening ​

Author: Ayushi Jolotia, Parikshit Pareek
Published: 7/15/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2607.12241v1 Announce Type: cross Abstract: Machine-learned surrogates for the AC power flow (ACPF) problem amortize the cost of repeated solves on a fixed network, but lose one to two orders of magnitude of accuracy when a line outage changes the topology. This degradation is an operator shif...

📖 Read original article


112. Rough Path Signature-Guided Geometry Augmentation for Few-Shot Industrial Surface Defect Detection ​

Author: Jiaqi Kuang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, math.PR

arXiv:2607.12245v1 Announce Type: cross Abstract: Few-shot industrial defect detection remains difficult for standard supervised detectors, which achieve poor performance on boundary-dominated industrial defects. This paper proposes rough path signature-guided geometry augmentation (RPS-GA), a geome...

📖 Read original article


113. When Directional Accuracy Lies: A Base-Rate-Honest Benchmark for LoRA-Adapted TimesFM on Equity Forecasting ​

Author: Taizhen Cheung, SA Kwon
Published: 7/15/2026, 4:00:00 AM
Categories: q-fin.ST, cs.LG

arXiv:2607.12248v1 Announce Type: cross Abstract: Large pretrained time-series models such as TimesFM are attractive for financial forecasting, but raw directional accuracy is a misleading scoreboard in equity markets. An early LoRA adapter in this project appeared to reach roughly 80% directional a...

📖 Read original article


114. On-Device Deep Research at 4B: Exposure Bounds Faithfulness, Retrieval Bounds Coverage ​

Author: Vinay Kumar Chaganti
Published: 7/15/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.IR, cs.LG

arXiv:2607.12257v1 Announce Type: cross Abstract: On-device research agents search a corpus, read sources, and write a cited brief on a personal laptop. Whether their citations are faithful, and at what cost, is unmeasured for a deployable small model. This study fixes one 4B generator on a 24 GB la...

📖 Read original article


115. Auditing Data Leakage in Whole-Slide Image Multimodal Benchmarks ​

Author: Wenhao Zhang, Zhongliang Zhou, John Kang, Sheng Li
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.12278v1 Announce Type: cross Abstract: Recent vision-language models (VLMs) for computational pathology report striking zero-shot performance on whole-slide image (WSI) visual question answering (VQA) benchmarks. We audit these claims and find them fundamentally compromised by data leakag...

📖 Read original article


116. A Shared Subcircuit Lets LLMs Count Down Across Tasks ​

Author: Jacob Dunefsky, Wes Gurnee, Emmanuel Ameisen
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.12279v1 Announce Type: cross Abstract: Writing a sentence of exactly twelve words; ending a DNA sequence at the right codon; formatting an ASCII table. These are all tasks that language models can do that requires tracking how many tokens remain before a target. In this work, we identify ...

📖 Read original article


117. SlimPer: Make Personalization Model Slim and Smart ​

Author: Siqi Wang, Xianjie Chen, Shaofeng Deng, Albert Chen, Romil Shah, Jiawei Huang, Zhaoqin Wang, Zhang Zhang, Yiqun Liu, Meilei Jiang, Anish Dubey, Moyan Mei, Tongxin Wang, Nathan Berrebbi, Misael Manjarres, Armand Sauzay, Shardul Kothapalli, Aryaman Vinchhi, Kevin Johnstone, Juheon Lee, Gufan Yin, Ziheng Huang, Justin Lin, Mert Terzihan, Yilin Qi, Cynthia Yang, Colin Peppler, Qi Ding, Ruohan Sun, Ge Song, Litao Deng, Parichay Kapoor, Matt Ma, Huihui Cheng, Jiyuan Zhang, Yanli Zhao, Yiping Han, Fangqiu Han, Ning Yao, Arun Singh, Jordan Edwards, Zhengyu Su, Abhishek Kumar, Guangdeng Liao, Ankit Asthana
Published: 7/15/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.12281v1 Announce Type: cross Abstract: Transformer-style architectures are increasingly adopted for industrial recommendation systems, yet they inherit a design premise misaligned with the task: generative models rely on per-token autoregressive prediction, which justifies maintaining lar...

📖 Read original article


118. The Sound of Absence: Audio-Language Embedding Models Struggle with Negation ​

Author: Chun-Yi Kuan, Hung-yi Lee
Published: 7/15/2026, 4:00:00 AM
Categories: eess.AS, cs.AI, cs.CL, cs.LG, cs.SD

arXiv:2607.12290v1 Announce Type: cross Abstract: Audio-language embedding models such as CLAP are widely evaluated on matching present sound events, but rarely on negation. We show this affirmation-only evaluation hides a key limitation: these models fail to encode negated sound concepts, mapping a...

📖 Read original article


119. What Does a Temporal Benchmark Score Measure? Decomposing Channel Use in Video VLM Evaluation ​

Author: Farrukh Rahman
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.12304v1 Announce Type: cross Abstract: A score on a temporal video question answering benchmark is meant to measure that a model has temporal understanding, but it conflates two questions. 1. The task question: is the question even temporal, does it need several frames and their order? an...

📖 Read original article


120. Robust Design of Integrated Sensing and Communication in LEO Satellite Systems ​

Author: Hezhen Yang, Xiaoming Chen, Qi Wang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT

arXiv:2607.12337v1 Announce Type: cross Abstract: With the growing demand for satellite sensing and communication, the limited wireless resources are difficult to support multiple satellite systems. Therefore, it is desired to investigate integrated sensing and communication (ISAC) in low Earth orbi...

📖 Read original article


121. Thompson Sampling Is 2-Competitive for Mistakes ​

Author: Mark Sellke, Gregory Valiant
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2607.12389v1 Announce Type: cross Abstract: We consider Bayesian bandit models and prove that Thompson sampling makes at most twice the expected number of mistakes (selections of a suboptimal arm) as any other policy. Our analysis applies as long as the latent arm processes are independent and...

📖 Read original article


122. MESH: Scaling Up Retrieval with Heterogeneous Content Unification ​

Author: Jiaxing Qu, Yilin Chen, Junpeng Hou, Jinfeng Rao, Olafur Gudmundsson, Sai Xiao, Huizhong Duan
Published: 7/15/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.12392v1 Announce Type: cross Abstract: Optimizing large-scale retrieval hinges on the ability to efficiently surface candidates across diverse content tiers. However, to capture segments such as fresh and long-tail content, modern systems typically resort to a fragmented "zoo" of speciali...

📖 Read original article


123. Language Identification with Succinct Machine-Independent Traces ​

Author: Moses Charikar, Jon Kleinberg, Chirag Pabbaraju
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.DS, cs.LG

arXiv:2607.12443v1 Announce Type: cross Abstract: Motivated by the power of large language models, there has been renewed interest in the Gold-Angluin model of language identification in the limit, with an eye toward variants of the model that might overcome the negative results for its original for...

📖 Read original article


124. Steering Diffusion Models via Class-Contrastive Influence for Few-Shot Medical Classification ​

Author: Jeeyung Kim, Erfan Esmaeili, Qiang Qiu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.12464v1 Announce Type: cross Abstract: When labeled data are scarce, off-the-shelf diffusion models can augment training sets for few-shot medical image classification, but not all generated samples are equally useful for the downstream task. Existing approaches largely improve synthetic ...

📖 Read original article


125. Mind the Gap: Promises and Pitfalls of Hierarchical Planning in LeWorldModel ​

Author: Niccol`o Caselli, Francesco Massafra, Samuele Punzo, Salvatore Lo Sardo, Ippokratis Pantelidis, Sathya Kamesh Bhethanabhotla
Published: 7/15/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.12547v2 Announce Type: cross Abstract: We investigate whether temporal hierarchy can improve LeWorldModel on long-horizon goal-conditioned control. We introduce Hi-LeWM, an extension that freezes the pretrained low-level LeWM and adds high-level planning over latent subgoals. We evaluate ...

📖 Read original article


126. Deep Learning-based Surrogate Modelling of the LOD Method for Multiscale Problems ​

Author: Marc Haltmayer, Jaemin Seo, Yuseung Lee, Sungyeop Lee, Jaehoon Jeong, Jae Yong Lee
Published: 7/15/2026, 4:00:00 AM
Categories: math.NA, cs.AI, cs.LG, cs.NA

arXiv:2607.12570v1 Announce Type: cross Abstract: Multiscale problems are notoriously difficult to tackle using traditional numerical methods, as accurately resolving fine-scale features often requires prohibitively fine discretizations. This challenge is particularly pronounced in applications such...

📖 Read original article


127. Environment Parameter Gradient Theorem for Policy-Environment Co-Design in Reinforcement Learning ​

Author: Amber Srivastava
Published: 7/15/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY

arXiv:2607.12590v1 Announce Type: cross Abstract: Reinforcement learning (RL) is traditionally concerned with learning a control policy for a fixed environment. In many engineering systems, however, the environment itself is alterable: physical or operational parameters can be tuned to shape the tra...

📖 Read original article


128. Gradient-free learning of a closed-loop wall controller for turbulent drag reduction ​

Author: Giorgio Maria Cavallazzi, Miguel P'erez Cuadrado, Alfredo Pinelli
Published: 7/15/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG

arXiv:2607.12626v1 Announce Type: cross Abstract: Closed-loop wall control learnt by multi-agent reinforcement learning can lower skin-friction drag in turbulent channels, but these gradient-based policies are trained on small periodic boxes and exhibit reduced performance when carried over to a lar...

📖 Read original article


129. Label-Decoupled Style Augmentation for Domain Generalization in Multi-Label Remote Sensing Scene Classification ​

Author: Alaa Almouradi, Erchan Aptoula
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.12704v2 Announce Type: cross Abstract: Multi-label classification assigns several co-occurring labels to each aerial scene, yet deployed models often encounter data distributions different from their training. Feature-statistics augmentation such as MixStyle, EFDMix, and correlated style ...

📖 Read original article


130. Physically Consistent Parameter Inference: Transparent Machine Learning Emulation in High Energy Physics and Cosmology ​

Author: Jorge Alda, Jacobo Asorey, Alejandro Mir, Siannah Pe~naranda
Published: 7/15/2026, 4:00:00 AM
Categories: hep-ph, cs.LG

arXiv:2607.12726v1 Announce Type: cross Abstract: Global fits in high energy physics and cosmology often face the challenge of exploring high-dimensional parameter spaces with computationally expensive or topologically complex likelihood functions. In this work, we present a Machine Learning framewo...

📖 Read original article


131. LLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with Elenchos ​

Author: Julius Steiglechner, Lucas Mahler, Gabriele Lohmann
Published: 7/15/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.12733v1 Announce Type: cross Abstract: Large language models (LLMs) excel at pattern recognition and text generation, but their capacity for abductive inference - inferring latent hypotheses that explain observed behavior - remains poorly understood. Here, we introduce Elenchos (named aft...

📖 Read original article


132. Learning-enabled Acceleration of Scenario-based Model Predictive Control ​

Author: Trinh Tran, Binh Nguyen, Truong X. Nghiem
Published: 7/15/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.SY, eess.SY

arXiv:2607.12775v1 Announce Type: cross Abstract: Scenario-based model predictive control (SBMPC) is a variant of model predictive control (MPC) that explicitly accounts for uncertainty by optimizing control actions over multiple predicted scenarios. However, its computational complexity increases r...

📖 Read original article


133. Directional Constraints for Efficient Exploration in Safe Reinforcement Learning ​

Author: Paolo Magliano, Puze Liu, Jan Peters, Davide Tateo, Raffaello Camoriano
Published: 7/15/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.12784v1 Announce Type: cross Abstract: Reinforcement Learning has revolutionized the landscape of robotic research, allowing robust learning of complex robotic skills in simulation. However, real-world deployment in open-ended environments requires strong safety guarantees to prevent dang...

📖 Read original article


134. ANGLE: Angular Neural Generative Learning via Engression ​

Author: Rajdeep Pathak, Archi Roy, Tanujit Chakraborty
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2607.12833v1 Announce Type: cross Abstract: Circular data, representing angles or directions, are frequently encountered in computer vision, biology, geology, and meteorology. Traditional regression targets the conditional mean, which is often geometrically misleading for circular responses un...

📖 Read original article


135. Reproducible Reservoir Computing with Thermally Driven Superparamagnets: Controlling Temperature Sensitivity ​

Author: Zhengfei Chen, Alex Welbourne, Matthew O. A. Ellis, Dan A. Allwood, Eleni Vasilaki, Thomas J. Hayward
Published: 7/15/2026, 4:00:00 AM
Categories: cs.ET, cond-mat.mes-hall, cs.AI, cs.LG

arXiv:2607.12840v1 Announce Type: cross Abstract: Unconventional computing systems must demonstrate robust performance under real-world environmental conditions to enable practical deployments. We have recently proposed superparamagnetic nanodot ensembles driven by strain-induced magnetoelectric cou...

📖 Read original article


136. Toward Localizing and Repairing Bias in Transformer Attention Heads ​

Author: Sigma Jahan
Published: 7/15/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2607.12863v1 Announce Type: cross Abstract: Transformer language models are increasingly used as software components, yet biased outputs remain difficult to localize and repair inside the model. Existing fairness testing and repair methods largely operate at the input-output or retraining leve...

📖 Read original article


137. Deep4ge: DNN Training Trajectories for Fault Detection and Diagnosis ​

Author: Sigma Jahan
Published: 7/15/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2607.12868v1 Announce Type: cross Abstract: Deep learning systems often fail due to subtle implementation faults that alter training behavior. Recent work has studied how to detect and diagnose such failures from changes observed across training epochs. However, the software engineering commun...

📖 Read original article


138. Accelerated Mixing Time of Randomized Hamiltonian Monte Carlo ​

Author: Siddharth Mitra, Vishwak Srinivasan, Xiuyuan Wang, Andre Wibisono
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.DS, cs.LG, math.PR, math.ST, stat.CO, stat.TH

arXiv:2607.12902v1 Announce Type: cross Abstract: We show the Randomized Hamiltonian Monte Carlo (RHMC) algorithm has accelerated mixing time guarantees for sampling from log-concave probability distributions. RHMC proceeds by repeatedly simulating the continuous-time Hamiltonian dynamics for some r...

📖 Read original article


139. LatentFlow: A General Framework for Conditioning Stochastic Processes ​

Author: Louis Sharrock, Lachlan Astfalck, Henry Moss
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2607.12922v1 Announce Type: cross Abstract: Stochastic-process models are, as a rule, far easier to simulate than to condition. Non-linear observations, non-Gaussian likelihoods, black-box information, and global constraints all induce intractable conditional laws, requiring bespoke, model-spe...

📖 Read original article


140. Robustness of Deep Learning Models for PV Power Forecasting under NWP Forecast Errors: A Spatiotemporal and Physically Interpretable Analysis ​

Author: Dandan Chen, Yan Zhao, Xuepeng Chen
Published: 7/15/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG

arXiv:2607.12954v1 Announce Type: cross Abstract: Engineering use of AI forecasting models requires not only high nominal accuracy but also predictable behavior under uncertain inputs. In photovoltaic (PV) forecasting, this requirement is especially challenging because numerical weather prediction (...

📖 Read original article


141. Form, Not Content? A Preregistered, Placebo-Controlled Evaluation of Learned Error-Conditioned Self-Repair Through Prompts and Weights in Frozen Small Code Models ​

Author: Mehmet Iscan
Published: 7/15/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2607.12962v1 Announce Type: cross Abstract: Frozen small code LLMs are deployed locally, yet the information guiding a retry after a failed attempt is still measured without placebo controls in the self-repair literature. We treat a failed program as a conjecture and an execution counterexampl...

📖 Read original article


142. Ensemble Controlled-Flow Filtering for Implicit Data Assimilation ​

Author: Zhuoyuan Li, Yue Zhao, Ming Li
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, cs.NA, math.NA, math.OC

arXiv:2607.12975v1 Announce Type: cross Abstract: Data assimilation estimates the state of a dynamical system from model forecasts and incoming observations. Many observation mechanisms, however, are many-to-one, implicit, non-smooth, or accessible only through simulation, and need not provide the r...

📖 Read original article


143. Watermark Forensics for Generative Models: An Information-Theoretic Perspective ​

Author: Xiaoyu Li, Zheng Gao, Xiaoyan Feng, Jiaojiao Jiang, Yulei Sui, Jiankun Hu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CR, cs.IT, cs.LG, math.IT

arXiv:2607.13003v1 Announce Type: cross Abstract: A watermark in a generative model's output is usually asked only whether a text is machine-made. The same mark can do more: attribute it to the user who produced it, extract a hidden payload, or localize the part that survives editing. These form a f...

📖 Read original article


144. A Shortcut to Statistically Steady-State Turbulence with Flow Matching ​

Author: Gianluca Galletti, Gerald Gutenbrunner, William Hornsby, Lorenzo Zanisi, Naomi Carey, Stanislas Pamela, Johannes Brandstetter, Fabian Paischer
Published: 7/15/2026, 4:00:00 AM
Categories: physics.plasm-ph, cs.LG

arXiv:2607.13022v1 Announce Type: cross Abstract: Many nonlinear physical systems exhibit an initial transient phase in which perturbations grow before nonlinear interactions lead to a statistically steady state. While this saturated regime is of primary interest, direct numerical simulations must r...

📖 Read original article


145. Propheticus: Machine Learning Framework for the Development of Predictive Models for Reliable and Secure Software ​

Author: Jo~ao R. Campos, Marco Vieira, Ernesto Costa
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:1809.01898v2 Announce Type: replace Abstract: The growing complexity of software calls for innovative solutions that support the deployment of reliable and secure software. Machine Learning (ML) has shown its applicability to various complex problems and is frequently used in the dependability...

📖 Read original article


146. Diversity-Enriched Option-Critic ​

Author: Anand Kamat, Doina Precup
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2011.02565v2 Announce Type: replace Abstract: Temporal abstraction allows reinforcement learning agents to represent knowledge and develop strategies over different temporal scales. The option-critic framework has been demonstrated to learn temporally extended actions, represented as options, ...

📖 Read original article


147. The Limits and Potentials of Local SGD for Distributed Heterogeneous Learning with Intermittent Communication ​

Author: Kumar Kshitij Patel, Margalit Glasgow, Ali Zindari, Lingxiao Wang, Sebastian U. Stich, Ziheng Cheng, Nirmit Joshi, Nathan Srebro
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, math.OC, stat.ML

arXiv:2405.11667v2 Announce Type: replace Abstract: Local SGD is a popular optimization method in distributed learning, often outperforming other algorithms in practice, including mini-batch SGD. Despite this success, theoretically proving the dominance of local SGD in settings with reasonable data ...

📖 Read original article


148. Active Exploration via Autoregressive Generation of Missing Data ​

Author: Tiffany Tianhui Cai, Hongseok Namkoong, Daniel Russo, Kelly W Zhang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2405.19466v5 Announce Type: replace Abstract: We pose uncertainty quantification and exploration in online decision-making as a problem of training and generation from an autoregressive sequence model, an area experiencing rapid innovation. Our approach rests on viewing uncertainty as arising ...

📖 Read original article


149. MolMiner: Toward Controllable, 3D-Aware, Fragment-Based Molecular Design ​

Author: Raul Ortega-Ochoa, Tejs Vegge, Jes Frellsen
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci

arXiv:2411.06608v3 Announce Type: replace Abstract: We introduce MolMiner, a fragment-based, geometry-aware, and order-agnostic autoregressive model for molecular design. MolMiner supports high-dimensional conditional control over twelve physicochemical and structural properties from partial specifi...

📖 Read original article


150. RAFP: Identifying LLM Lineages via Rare-Region Fingerprints ​

Author: Yun-Yun Tsai, Jia Hao Liang, Chuan Guo, Junfeng Yang, Laurens van der Maaten
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2505.12682v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly released under restricted licenses, creating a growing need for robust model ownership verification. Existing fingerprinting methods are often fragile under downstream finetuning, require invasive train...

📖 Read original article


151. Kernel PCA for Out-of-Distribution Detection: Non-Linear Kernel Selection and Approximation ​

Author: Kun Fang, Qinghua Tao, Mingzhen He, Kexin Lv, Runze Yang, Haibo Hu, Xiaolin Huang, Jie Yang, Longbing Cao
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2505.15284v2 Announce Type: replace Abstract: Out-of-Distribution (OoD) detection is vital for the reliability of deep neural networks, the key of which lies in effectively characterizing the disparities between OoD and In-Distribution (InD) data. In this work, such disparities are exploited t...

📖 Read original article


152. Inclusive Federated Learning Through Compliance-Weighted Noise Allocation in Healthcare AI ​

Author: Santhosh Parampottupadam, Melih Co\c{s}\u{g}un, Sarthak Pati, Maximilian Zenk, Saikat Roy, Dimitrios Bounias, Benjamin Hamm, Sinem Sav, Ralf Floca, Klaus Maier-Hein
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR, cs.DC

arXiv:2505.22108v4 Announce Type: replace Abstract: Background: Federated learning (FL) enables collaborative training of clinical AI models without centralizing patient data, but adoption is limited by privacy concerns, heterogeneous institutional compliance, and resource disparities; standard diff...

📖 Read original article


153. Diagrams-to-Dynamics (D2D): Exploring Causal Loop Diagram Leverage Points under Uncertainty ​

Author: Jeroen F. Uleman, Loes Crielaard, Leonie K. Elsenburg, Guido A. Veldhuis, Naja Hulvej Rod, Rick Quax, V'itor V. Vasconcelos
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2508.05659v4 Announce Type: replace Abstract: Background: Causal loop diagrams (CLDs) are widely used in health and environmental research to represent hypothesized causal structures underlying complex problems. However, as qualitative and static representations, CLDs are limited in their abil...

📖 Read original article


154. Efficient Conformal Prediction for Regression Models under Label Noise ​

Author: Yahav Cohen, Jacob Goldberger, Tom Tirer
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.15120v2 Announce Type: replace Abstract: In high-stakes scenarios, such as medical imaging applications, it is critical to equip the predictions of a regression model with reliable confidence intervals. Recently, Conformal Prediction (CP) has emerged as a powerful statistical framework th...

📖 Read original article


155. Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning ​

Author: Xin Qiu, Yulu Gan, Conor F. Hayes, Qiyao Liang, Yinggan Xu, Roberto Dailey, Elliot Meyerson, Babak Hodjat, Risto Miikkulainen
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE

arXiv:2509.24372v3 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) for downstream tasks is an essential stage of modern AI deployment. Reinforcement learning (RL) has emerged as the dominant fine-tuning paradigm, underpinning many state-of-the-art LLMs. In contrast, evoluti...

📖 Read original article


156. Variational Mixture of Graph Neural Experts for Alzheimer's Disease Recognition across Frequency Bands in EEG Brain Networks ​

Author: Jun-En Ding, Anna Zilverstand, Shihao Yang, Albert Chih-Chieh Yang, Feng Liu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.11917v3 Announce Type: replace Abstract: Dementia disorders such as Alzheimer's disease (AD) and frontotemporal dementia (FTD) exhibit overlapping electrophysiological signatures in EEG that challenge accurate diagnosis. Existing EEG-based methods are limited by full-band frequency analys...

📖 Read original article


157. Higher Embedding Dimension Creates a Stronger World Model for a Simple Sorting Task ​

Author: Brady Bhalla, Honglu Fan, Nancy Chen, Tony Yue YU
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.18315v2 Announce Type: replace Abstract: We investigate how embedding dimension affects the emergence of an internal "world model" in a transformer trained with reinforcement learning to perform bubble-sort-style adjacent swaps. Models achieve high accuracy even with very small embedding ...

📖 Read original article


158. When and Why Does Multi-Agent Debate Fail and Does It Really Underperform? ​

Author: Yongqiang Chen, Gang Niu, James Cheng, Bo Han, Masashi Sugiyama
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.20963v2 Announce Type: replace Abstract: Multi-agent debate (MAD) was proposed as a promising approach for ensembling the wisdom of multiple large language models (LLMs) to improve reasoning and provide effective supervision to superhuman LLMs. However, increasing empirical evidence sugge...

📖 Read original article


159. Auditable Context-Aware HFMD Forecasting with Structured LLM Agents ​

Author: Joongwon Chae, Runming Wang, Chen Xiong, Gong Yunhan, Lian Zhang, Ji Jiansong, Dongmei Yu, Peiwu Qin
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.MA

arXiv:2511.23276v2 Announce Type: replace Abstract: Effective HFMD surveillance requires forecasts capturing both time-series patterns and contextual drivers such as school calendars, weather, and policy or surveillance reports. In clinical settings, forecasts must be trusted and actionable; thus, b...

📖 Read original article


160. Efficiently Learning Branching Networks for Multitask Algorithmic Reasoning ​

Author: Dongyue Li, Zhenshuo Zhang, Minxuan Duan, Edgar Dobriban, Hongyang R. Zhang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DS

arXiv:2512.01113v2 Announce Type: replace Abstract: Algorithmic reasoning -- the ability to perform step-by-step logical inference -- is a synthetic benchmark for evaluating multi-step reasoning abilities, designed for graph neural networks and also for transformer models. Prior work has evaluated r...

📖 Read original article


161. AnySleep: a channel-agnostic deep learning system for high-resolution sleep staging in multi-center cohorts ​

Author: Niklas Grieger, Jannik Raskob, Siamak Mehrkanoon, Stephan Bialonski
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, eess.SP, q-bio.QM

arXiv:2512.14461v2 Announce Type: replace Abstract: Sleep is essential for health, yet studying its dynamics requires manual sleep staging, a labor-intensive step in research and clinical care. Across centers, polysomnography (PSG) recordings are traditionally scored in 30-s epochs for pragmatic, no...

📖 Read original article


162. Visual Species Recognition with Large Multimodal Models as Post-Hoc Correctors ​

Author: Tian Liu, Anwesha Basu, James Caverlee, Shu Kong
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2512.15748v2 Announce Type: replace Abstract: Visual Species Recognition (VSR) is a fundamental task in scientific disciplines that require species-level identification, including ecology, palynology, evolutionary biology, systematics, and phylogenetics. Automating VSR through machine learning...

📖 Read original article


163. Calibratable Disambiguation Loss for Multi-Instance Partial-Label Learning ​

Author: Wei Tang, Yin-Fang Yang, Weijia Zhang, Min-Ling Zhang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.17788v2 Announce Type: replace Abstract: Multi-instance partial-label learning (MIPL) is a weakly supervised framework that extends the principles of multi-instance learning (MIL) and partial-label learning (PLL) to address the challenges of inexact supervision in both instance and label ...

📖 Read original article


164. Mechanistic Evidence for Preserved-but-Misaligned Representations in Non-IID FedAvg ​

Author: Muhammad Haseeb, Salaar Masood, Muhammad Abdullah Sohail, Mohammad Fatim Shoaib, Muhammad Tahir
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.23043v3 Announce Type: replace Abstract: Federated Averaging (FedAvg) often degrades under non-IID client data, but it remains unclear whether this degradation reflects the loss of client-learned representations or a failure to use representations that are still present. We study this que...

📖 Read original article


165. Graph Regularized PCA ​

Author: Antonio Briola, Marwin Schmidt, Fabio Caccioli, Carlos Ros Perez, James Singleton, Christian Michler, Tomaso Aste
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.10199v2 Announce Type: replace Abstract: Multivariate data often exhibit complex dependencies that violate the assumption of isotropic residual noise. For such cases, we introduce Graph Regularized PCA (GR-PCA). It is a graph-based regularization of PCA that incorporates the dependency st...

📖 Read original article


166. Tracking Drift: Variation-Aware Entropy Scheduling for Non-Stationary Reinforcement Learning ​

Author: Tongxi Wang, Zhuoyang Xia, Xinran Chen, Shan Liu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.19624v3 Announce Type: replace Abstract: Real-world reinforcement learning often faces environment drift, but most existing methods rely on static entropy coefficients/target entropy, causing over-exploration during stable periods and under-exploration after drift, and leaving unanswered ...

📖 Read original article


167. Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models ​

Author: Hao Wang, Hao Gu, Hongming Piao, Kaixiong Gong, Yuxiao Ye, Xiangyu Yue, Sirui Han, Yike Guo, Dapeng Wu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2602.02244v3 Announce Type: replace Abstract: The standard post-training recipe for large reasoning models, supervised fine-tuning followed by reinforcement learning (SFT-then-RL), may limit the benefits of the RL stage: while SFT imitates expert demonstrations, it often causes overconfidence ...

📖 Read original article


168. A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs ​

Author: Zijie Liu, Jie Peng, Jinhao Duan, Zirui Liu, Kaixiong Zhou, Mingfu Liang, Luke Simon, Xi Liu, Zhaozhuo Xu, Tianlong Chen
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.19938v2 Announce Type: replace Abstract: Sparse Mixture-of-Experts (SMoE) architectures are increasingly used to scale large language models efficiently, delivering strong accuracy under fixed compute budgets. However, SMoE models often suffer from severe load imbalance across experts, wh...

📖 Read original article


169. Understanding LoRA as Knowledge Memory: An Empirical Analysis ​

Author: Seungju Back, Dongwoo Lee, Naun Kang, Taehee Lee, S. K. Hong, Youngjune Gwon, Sungjin Ahn
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.01097v4 Announce Type: replace Abstract: Continuous knowledge updating for pre-trained large language models (LLMs) is increasingly necessary yet remains challenging. Although inference-time methods like In-Context Learning (ICL) and Retrieval-Augmented Generation (RAG) are popular, they ...

📖 Read original article


170. Same Compression Principle, Different Geometry: Rate-Distortion Signatures Dissociate Biological and Artificial Visual Systems ​

Author: Leyla Roksan Caglar, Pedro A. M. Mediano, Baihan Lin
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.IT, math.IT, q-bio.NC

arXiv:2603.01568v2 Announce Type: replace Abstract: Efficient coding theory predicts that biological perceptual systems compress sensory input optimally under resource constraints, with the systematic structure of errors reflecting the geometry of that compression. Here we operationalize this princi...

📖 Read original article


171. MUSA-PINN: Multi-scale Weak-form Physics-Informed Neural Networks for Fluid Flow in Complex Geometries ​

Author: Weizheng Zhang, Xunjie Xie, Hao Pan, Xiaowei Duan, Bingteng Sun, Qiang Du, Lin Lu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.08465v3 Announce Type: replace Abstract: While Physics-Informed Neural Networks (PINNs) offer a mesh-free approach to solving fluid-flow PDEs, standard point-wise residual minimization suffers from convergence pathologies in topologically complex domains like Triply Periodic Minimal Surfa...

📖 Read original article


172. Generalization and Memorization in Rectified Flow ​

Author: Mingxing Rao, Daniel Moyer
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2603.13421v2 Announce Type: replace Abstract: Generative models based on the Flow Matching objective, particularly Rectified Flow, have emerged as a dominant paradigm for efficient, high-fidelity image synthesis. However, while existing research heavily prioritizes generation quality and archi...

📖 Read original article


173. High-Dimensional Gaussian Mean Estimation under Realizable Contamination ​

Author: Ilias Diakonikolas, Daniel M. Kane, Thanasis Pittas
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.DS, math.ST, stat.ML, stat.TH

arXiv:2603.16798v2 Announce Type: replace Abstract: We study mean estimation for a Gaussian distribution with identity covariance in $\mathbb{R}^d$ under a missing data scheme termed realizable $\epsilon$-contamination model. In this model an adversary can choose a function $r(x)$ between 0 and $\ep...

📖 Read original article


174. ScatterPrism: convergence for generative simulation and inverse problems in particle and nuclear physics ​

Author: Zeyu Xia, Tyler Kim, Trevor Reed, Judy Fox, Geoffrey Fox, Adam Szczepaniak
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, nucl-ex, physics.data-an, physics.ins-det

arXiv:2604.01313v3 Announce Type: replace Abstract: High-fidelity simulations and complex inverse problems, such as detector modeling and unfolding, are computationally intensive bottlenecks across subatomic physics, yet essential for accurate physical interpretation. While Conditional Flow Matching...

📖 Read original article


175. Beyond Logit Adjustment: A Residual Decomposition Framework for Long-Tailed Reranking ​

Author: Zhanliang Wang, Hongzhuo Chen, Quan Minh Nguyen, Mian Umair Ahsan, Kai Wang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.01506v2 Announce Type: replace Abstract: Long-tailed classification, where a small number of frequent classes dominate many rare ones, remains challenging because models systematically favor frequent classes at inference time. Existing post-hoc methods such as logit adjustment address thi...

📖 Read original article


176. Learning an Interpretable Risk Scoring System for Maximizing Decision Net Benefit ​

Author: Wenhao Chi, \c{S}. .Ilker Birbil
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2604.04241v2 Announce Type: replace Abstract: Risk scoring systems are widely used in high-stakes domains to assist decision-making. However, existing approaches often focus on optimizing predictive accuracy or likelihood-based criteria, which may not align with the main goal of maximizing uti...

📖 Read original article


177. NetForge RL: A Multi-Agent Simulation Environment for Cyber Defense with Durative Actions ​

Author: Igor Jankowski
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.MA

arXiv:2604.09523v3 Announce Type: replace Abstract: Training reinforcement-learning agents for cyber defense requires an environment that reflects the operational setting: noisy, partial observations, several defenders coordinating across a network, and an adaptive adversary realized through self-pl...

📖 Read original article


178. Neuro-Symbolic ODE Discovery with Latent Grammar Flow ​

Author: Karin Yu, Eleni Chatzi, Georgios Kissas
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CE, cs.SC

arXiv:2604.16232v2 Announce Type: replace Abstract: Understanding natural and engineered systems often relies on symbolic formulations, such as differential equations, which provide interpretability and transferability beyond black-box models. We introduce Latent Grammar Flow (LGF), a neuro-symbolic...

📖 Read original article


179. PRIME: Protein Representation via Physics-Informed Multiscale Equivariant Hierarchies ​

Author: Viet Thanh Duy Nguyen, John K. Johnstone, Truong-Son Hy
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.01625v3 Announce Type: replace Abstract: Proteins are inherently multiscale physical systems whose functional properties emerge from coordinated structural organization across multiple spatial resolutions, ranging from atomic interactions to global fold topology. However, existing protein...

📖 Read original article


180. Inference-Time Machine Unlearning via Gated Activation Redirection ​

Author: Vin'icius Conte Turani, Ot'avio Parraga, Jo~ao Vitor Boer Abitante, Kristen K. Arguello, Joana Pasquali, Ramiro N. Barros, Flavio du Pin Calmon, Christian Mattjie, Rodrigo C. Barros, Lucas S. Kupssinsk"u
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.12765v3 Announce Type: replace Abstract: Large Language Models memorize vast amounts of training data, raising concerns regarding privacy, copyright infringement, and safety. Machine unlearning seeks to remove the influence of a targeted forget set while preserving model performance, idea...

📖 Read original article


181. A Geometric Approach to Constrained Online Learning ​

Author: Dhruv Sarkar, Abhishek Sinha
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2605.21107v2 Announce Type: replace Abstract: We study constrained online convex optimization with adversarial time-varying constraints. At each round the learner acts before observing the loss and constraint, and is compared with the best fixed action satisfying all constraints in hindsight. ...

📖 Read original article


182. AMUSE: Anytime Muon with Stable Gradient Evaluation ​

Author: Jueun Kim, Baekrok Shin, Jihun Yun, Beomhan Baek, Minhak Song, Chulhee Yun
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.22432v2 Announce Type: replace Abstract: Modern deep learning commonly relies on AdamW with prescribed learning rate schedules, but recent works challenge both components: Schedule-Free optimization removes explicit schedules via iterate averaging, and Muon improves the update geometry by...

📖 Read original article


183. Hierarchical Synthetic Tabular Data Generation: A Hybrid Top-Down and Bottom-Up Framework ​

Author: Junfeng Nie, Alvin Jin, Xiaohui Chen
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.28198v2 Announce Type: replace Abstract: Existing approaches for synthetic tabular data generation are based on either purely generative models or LLMs, both of which struggle with data heterogeneity, logical consistency, rare-event coverage, and robustness in low-data regimes. In this pa...

📖 Read original article


184. dMX: Differentiable Mixed-Precision Assignment for Low-Precision Floating-Point Formats ​

Author: Giuseppe Franco, Ian Colbert, Pablo Monteagudo-Lago, Felix Marty, Nicholas Fraser
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.04115v2 Announce Type: replace Abstract: Quantizing large language models (LLMs) to low-precision floating-point representations is central to efficient deployment, yet applying a single bit-width uniformly across all layers is sub-optimal in terms of both performance and accuracy. This w...

📖 Read original article


185. Rethinking Incompleteness: Formalizing Protocol Divergence and Train-Once Learning for Robust IMVC ​

Author: Haolu Liu, Xiyue Wang, Xuanting Xie, Liangjian Wen, Zhao Kang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.04857v2 Announce Type: replace Abstract: Standard IMVC evaluation retrains separate models for different missing-data configurations. We show that this paradigm obscures a fundamental vulnerability: missing rate alone is insufficient to characterize data incompleteness. Specifically, we s...

📖 Read original article


186. Explaining Data Mixing Scaling Laws ​

Author: Rui Dai, Shuran Zheng
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.08167v2 Announce Type: replace Abstract: Recent research has established empirical scaling laws to predict model performance on multi-domain data mixtures. However, a theoretical understanding of these model loss behaviors remains absent. In this work, we propose a unified framework to ex...

📖 Read original article


187. Keep Policy Gradient in Charge: Sibling-Guided Credit Distillation for Long-Horizon Tool-Use Agents ​

Author: Tianyu Ding, Jianhong Xin, Juan Pablo De la Cruz Weinstein
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2606.12634v3 Announce Type: replace Abstract: Long-horizon tool-use reinforcement learning learns from outcome verification, but trajectory-level advantages are broadcast over reasoning, API, and answer tokens. Direct self-distillation can supply a denser signal, but in our experiments it can ...

📖 Read original article


188. Some Complexity Results for Robustness Verification for Binarized Neural Networks ​

Author: Harshit Goyal, Sudakshina Dutta
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.CC

arXiv:2606.18918v3 Announce Type: replace Abstract: This paper studies the computational complexity of verification problems for Binarized Neural Networks (BNNs), where activations (and sometimes weights) are binary. We analyze two problems: satisfiability and robustness under uniform image occlusio...

📖 Read original article


189. Prime Fourier Embeddings: A Principled Basis for Modular Arithmetic ​

Author: Hyunsang Hwang, Suhyun Bae, Donghun Lee
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.23044v2 Announce Type: replace Abstract: Numbers have algebraic structure that standard neural embeddings often fail to expose. We introduce Prime Fourier Embeddings (PFE), which encode integers as prime-indexed (cos, sin) pairs derived from the harmonic analysis of Q, providing a pre-str...

📖 Read original article


190. Explaining Temporal Graph Neural Networks via Feature-induced Information Flow ​

Author: Ping Xiong, Thomas Schnake, Klaus-Robert M"uller, Shinichi Nakajima
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.27201v2 Announce Type: replace Abstract: Event-based Temporal Graph Neural Networks (ETGNNs) have demonstrated strong performance across a wide range of applications, including social network analysis, epidemic tracing, recommender systems, and political event forecasting. However, their ...

📖 Read original article


191. Beyond the Hard Budget: Sparsity Regularizers for More Interpretable Top-k Sparse Autoencoders ​

Author: Nathana"el Jacquier, Maria Vakalopoulou, Mahdi S. Hosseini
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.27321v2 Announce Type: replace Abstract: Sparse autoencoders (SAEs) have become a leading tool for interpreting the representations of vision foundation models, decomposing their polysemantic activations into a larger set of sparse, more monosemantic features. The Top-$k$ SAE, a now-stand...

📖 Read original article


192. Structure-Specific Representational Priors Causally Control the Grokking Delay ​

Author: Gunner Levi Howe
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.04333v3 Announce Type: replace Abstract: Grokking -- generalization long after training-set interpolation -- has been accelerated by structure-agnostic interventions (gradient filtering, weight-norm clamping, geometric penalties). Whether the delay specifically measures the time to form t...

📖 Read original article


193. MLPTR-CC: Multi-label Pathology Test Recommendation using Classifier Chains and SHAP ​

Author: Abu Rafe Md Jamil, Nayan Malakar
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.08299v2 Announce Type: replace Abstract: Diagnostic decision making often relies on a sequence of pathology tests that bridge patient symptoms and final disease diagnosis. Existing clinical decision-support systems typically focus on predicting single diseases and do not explicitly recomm...

📖 Read original article


194. Learning in Curved Weight Space:Exponential-Linear Weight Reparameterization for Improved Optimization ​

Author: Ethan Smith
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.09967v2 Announce Type: replace Abstract: Many neural networks operations have a multiplicative nature rather than additive: halving or doubling a norm are analogous relatively but require unequal optimization distances when taking linear steps. Adaptive optimizers such as Adam normalize u...

📖 Read original article


195. Adaptive Compute in Latent World Models: When Depth Helps, Hurts, or Doesn't Matter ​

Author: Achyuthan Sivasankar
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.10203v2 Announce Type: replace Abstract: Adaptive-compute world models -- early-exit or mixture-of-depths predictors that spend variable depth per step -- assume depth buys better predictions and can be routed adaptively. In autoregressive rollouts, the first assumption requires depth's p...

📖 Read original article


196. PRISM Edit: One Vector for All Temporal Answers ​

Author: Chen Huang, Qi Zheng, Ruiqin Zheng, Long Zeng, Yuantong Xu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11327v2 Announce Type: replace Abstract: Model editing keeps large language models (LLMs) up to date without retraining, but temporal facts expose a limitation of the prevailing locate-and-edit paradigm: an update is not always a replacement. When a fact changes, the new answer should bec...

📖 Read original article


197. Invariant Learning Dynamics of Transformers in Inductive Reasoning Tasks ​

Author: Tiberiu Musat, Tiago Pimentel, Nicolas Zucchet, Thomas Hofmann
Published: 7/15/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11875v2 Announce Type: replace Abstract: We present a theoretical framework to explain the emergence of inductive reasoning abilities in Transformer language models. While previous works on Transformer learning dynamics have so far been mostly tied to specific tasks, we study a generalize...

📖 Read original article


198. Spectral Diffusion Processes ​

Author: Angus Phillips, Thomas Seror, Michael Hutchinson, Valentin De Bortoli, Arnaud Doucet, Emile Mathieu
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2209.14125v3 Announce Type: replace-cross Abstract: Diffusion models have proven to be a flexible and effective framework for modelling probability distributions on finite-dimensional spaces. However, many physical modelling problems such as time series are naturally described over function sp...

📖 Read original article


199. The TopCoW Challenge -- Topology-Aware Circle of Willis Segmentation for CT and MR Angiography ​

Author: Kaiyuan Yang, Fabio Musio, Yihui Ma, Norman Juchler, Johannes C. Paetzold, Rami Al-Maskari, Luciano H"oher, Hongwei Bran Li, Ibrahim Ethem Hamamci, Anjany Sekuboyina, Suprosanna Shit, Houjing Huang, Chinmay Prabhakar, Ezequiel de la Rosa, Bastian Wittmann, Diana Waldmannstetter, Florian Kofler, Fernando Navarro, Martin J. Menten, Ivan Ezhov, Daniel Rueckert, Iris N. Vos, Ynte M. Ruigrok, Birgitta K. Velthuis, Hugo J. Kuijf, Pengcheng Shi, Wei Liu, Ting Ma, Maximilian R. Rokuss, Yannick Kirchhoff, Fabian Isensee, Klaus Maier-Hein, Chengcheng Zhu, Huilin Zhao, Philippe Bijlenga, Julien H"ammerli, Catherine Wurster, Laura Westphal, Jeroen Bisschop, Elisa Colombo, Hakim Baazaoui, Hannah-Lea Handelsmann, Andrew Makmur, James Hallinan, Amrish Soundararajan, Benedikt Wiestler, Jan S. Kirschke, Evamaria O. Riedel, Roland Wiest, Emmanuel Montagnon, Laurent Letourneau-Guillon, Kwanseok Oh, Dahye Lee, Orhun Utku Aydin, Adam Hilbert, Jana Rieger, Dimitrios Rallios, Satoru Tanioka, Alexander Koch, Dietmar Frey, Abdul Qayyum, Moona Mazher, Steven Niederer, Nico Disch, Julius C. Holzschuh, Dominic LaBella, Francesco Galati, Daniele Falcetta, Maria A. Zuluaga, Chaolong Lin, Haoran Zhao, Zehan Zhang, Minghui Zhang, Xin You, Hanxiao Zhang, Guang-Zhong Yang, Yun Gu, Sinyoung Ra, Jongyun Hwang, Hyunjin Park, Junqiang Chen, Marek Wodzinski, Henning M"uller, Nesrin Mansouri, Florent Autrusseau, Cansu Yalcin, Rachika E. Hamadache, Clara Lisazo, Joaquim Salvi, Adri`a Casamitjana, Xavier Llad'o, Uma Maria Lal-Trehan Estrada, Valeriia Abramova, Luca Giancardo, Arnau Oliver, Paula Casademunt, Adrian Galdran, Matteo Delucchi, Oscar Camara, Jialu Liu, Haibin Huang, Yue Cui, Zehang Lin, Yusheng Liu, Shunzhi Zhu, Tatsat R. Patel, Adnan H. Siddiqui, Vincent M. Tutino, Maysam Orouskhani, Huayu Wang, Mahmud Mossa-Basha, Yuki Sato, Sven Hirsch, Susanne Wegener, Bjoern Menze
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, q-bio.QM, q-bio.TO

arXiv:2312.17670v5 Announce Type: replace-cross Abstract: The Circle of Willis (CoW) is an important network of arteries connecting major circulations of the brain. Its vascular architecture is believed to influence the risk, severity, and outcome of serious neurovascular diseases. However, characte...

📖 Read original article


200. Beyond Perceptual Distance: Discrepancy Assessment on Deep Representation for Out-of-Distribution Detection with Diffusion Model ​

Author: Kun Fang, Zuopeng Yang, Haibo Hu, Xiaolin Huang, Jie Yang, Qinghua Tao
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2409.10094v3 Announce Type: replace-cross Abstract: Out-of-Distribution (OoD) detection aims to justify whether a given sample is from the training distribution of the classifier-under-protection, i.e., In-Distribution (InD), or from an unknown out distribution. Recent researches have leverage...

📖 Read original article


201. A Comprehensive Evaluation of Deep Learning Object Detection Models on Heterogeneous Edge Devices ​

Author: Daghash K. Alqahtani, Muhammad Aamir Cheema, Maria A. Rodriguez, Adel N. Toosi
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.AR, cs.DC, cs.LG, cs.SE

arXiv:2409.16808v3 Announce Type: replace-cross Abstract: Modern applications such as autonomous vehicles, intelligent surveillance, and smart city systems increasingly require object detection on resource-constrained edge devices. Yet, there is still limited understanding of how different object de...

📖 Read original article


202. Enabling Energy-Efficient Simultaneous Multi-Task Reinforcement Learning through Spiking Neural Networks with Active Dendrites for Bio-inspired Generalist Agents ​

Author: Rachmad Vidya Wicaksana Putra, Avaneesh Devkota, Muhammad Shafique
Published: 7/15/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG

arXiv:2412.04847v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has demonstrated remarkable capabilities in training agents to solve complex tasks autonomously, such as mobile robots, UAVs/UGVs, and game-playing agents). However, scaling RL to master multiple tasks simultaneous...

📖 Read original article


203. Optimal and Provable Calibration in High-Dimensional Binary Classification: Angular Calibration and Platt Scaling ​

Author: Yufan Li, Pragya Sur
Published: 7/15/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.ME, stat.ML, stat.TH

arXiv:2502.15131v5 Announce Type: replace-cross Abstract: We study the fundamental problem of calibrating a linear binary classifier of the form $\sigma(\hat{w}^\top x)$, where the feature vector $x$ is Gaussian, $\sigma$ is a link function, and $\hat{w}$ is an estimator of the true linear weight $w...

📖 Read original article


204. Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence ​

Author: Qi Feng, Gu Wang
Published: 7/15/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2506.08121v2 Announce Type: replace-cross Abstract: We introduce a continuous policy-value iteration algorithm where the approximations of the value function of a stochastic control problem and the optimal control are simultaneously updated through Langevin-type dynamics. This framework applie...

📖 Read original article


205. Stochastic Quantum Spiking Neural Networks with Quantum Memory and Local Learning ​

Author: Jiechen Chen, Bipin Rajendran, Osvaldo Simeone
Published: 7/15/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2506.21324v3 Announce Type: replace-cross Abstract: Neuromorphic and quantum computing have recently emerged as promising paradigms for advancing artificial intelligence, each offering complementary strengths. Neuromorphic systems built on spiking neurons excel at processing time series data e...

📖 Read original article


206. Diffusion Denoiser-Aided Gyrocompassing ​

Author: Gershy Ben-Arie, Daniel Engelsman, Rotem Dror, Itzik Klein
Published: 7/15/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2507.21245v2 Announce Type: replace-cross Abstract: An accurate initial heading angle is essential for efficient and safe navigation across diverse domains. Unlike magnetometers, gyroscopes can provide accurate heading reference independent of the magnetic disturbances in a process known as gy...

📖 Read original article


207. Linear Regression under Missing or Corrupted Coordinates ​

Author: Ilias Diakonikolas, Jelena Diakonikolas, Daniel M. Kane, Jasper C. H. Lee, Thanasis Pittas
Published: 7/15/2026, 4:00:00 AM
Categories: cs.DS, cs.LG, math.ST, stat.ML, stat.TH

arXiv:2509.19242v2 Announce Type: replace-cross Abstract: We study multivariate linear regression under Gaussian covariates in two settings, where data may be erased or corrupted by an adversary under a coordinate-wise budget. In the incomplete data setting, an adversary may inspect the dataset and ...

📖 Read original article


208. Learning Latent Energy-Based Models via Interacting Particle Langevin Dynamics ​

Author: Joanna Marks, Tim Y. J. Wang, O. Deniz Akyildiz
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.CO

arXiv:2510.12311v2 Announce Type: replace-cross Abstract: We develop interacting particle algorithms for learning latent variable models with energy-based priors. To do so, we leverage recent developments in particle-based methods for solving maximum marginal likelihood estimation (MMLE) problems. S...

📖 Read original article


209. Hyperellipsoid Density Sampling: Exploitative Sequences to Accelerate High-Dimensional Numerical Optimization ​

Author: Julian G. Soltes
Published: 7/15/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, cs.NE

arXiv:2511.07836v5 Announce Type: replace-cross Abstract: The curse of dimensionality remains a persistent challenge in modern optimization problems. Expanding the search space into higher dimensions exponentiates the difficulty of finding optimal solutions, rendering traditional algorithms ineffici...

📖 Read original article


210. A Neurosymbolic Approach to Natural Language Formalization and Verification ​

Author: Chenyang An, Sam Bayless, Stefano Buliani, Darion Cassel, Byron Cook, Duncan Clough, R'emi Delmas, Nafi Diallo, Ferhat Erata, Nick Feng, Dimitra Giannakopoulou, Aman Goel, Aditya Gokhale, Joe Hendrix, Victor Heorhiadi, Marc Hudak, Dejan Jovanovi'c, Andrew M. Kent, Benjamin Kiesl-Reiter, Jeffrey J. Kuna, Nadia Labai, Joseph Lilien, Divya Raghunathan, Zvonimir Rakamari'c, Niloofar Razavi, Michael Tautschnig, Ali Torkamani, Nathaniel Weir, Michael W. Whalen, Jianan Yao
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG, cs.LO

arXiv:2511.09008v2 Announce Type: replace-cross Abstract: Large Language Models perform well at natural language interpretation and reasoning, but their lack of formal correctness guarantees limits their adoption in regulated industries like finance and health-care that operate under strict policies...

📖 Read original article


211. Learning and Testing Convex Functions ​

Author: Renato Ferreira Pinto Jr., Cassandra Marcussen, Elchanan Mossel, Shivam Nadimpalli
Published: 7/15/2026, 4:00:00 AM
Categories: cs.DS, cs.LG

arXiv:2511.11498v2 Announce Type: replace-cross Abstract: We consider the problems of \emph{learning} and \emph{testing} real-valued convex functions over Gaussian space. Despite the extensive study of function convexity across mathematics, statistics, and computer science, its learnability and test...

📖 Read original article


212. CROCS: A Two-Stage Clustering Framework for Behaviour-Centric Consumer Segmentation with Smart Meter Data ​

Author: Luke W. Yerbury, Ricardo J. G. B. Campello, G. C. Livingston Jr, Mark Goldsworthy, Lachlan O'Neil
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2601.10494v3 Announce Type: replace-cross Abstract: With grid operators confronting rising uncertainty from renewable integration and a broader push toward electrification, Demand-Side Management (DSM) -- particularly Demand Response (DR) -- has attracted significant attention as a cost-effect...

📖 Read original article


213. PluRel: Synthetic Data unlocks Scaling Laws for Relational Foundation Models ​

Author: Vignesh Kothapalli, Rishabh Ranjan, Valter Hudovernik, Vijay Prakash Dwivedi, Johannes Hoffart, Carlos Guestrin, Jure Leskovec
Published: 7/15/2026, 4:00:00 AM
Categories: cs.DB, cs.AI, cs.LG

arXiv:2602.04029v2 Announce Type: replace-cross Abstract: Relational Foundation Models (RFMs) facilitate data-driven decision-making by learning from complex multi-table databases. However, the diverse relational databases needed to train such models are rarely public due to privacy constraints. Whi...

📖 Read original article


214. Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making ​

Author: Ruoyu Chen, Shangquan Sun, Xiaoqing Guo, Sanyi Zhang, Kangwei Liu, Shiming Liu, Zhangcheng Wang, Qunli Zhang, Wei Wang, Hua Zhang, Xiaochun Cao
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2602.07008v3 Announce Type: replace-cross Abstract: Reliable models should not only predict correctly, but also justify decisions with acceptable evidence. Yet conventional supervised learning typically provides only class-level labels, allowing models to achieve high accuracy through shortcut...

📖 Read original article


215. Mobility-Aware Cache Framework for Scalable LLM-Based Human Mobility Simulation ​

Author: Hua Yan, Heng Tan, Yingxue Zhang, Yu Yang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2602.16727v2 Announce Type: replace-cross Abstract: Simulating large-scale human mobility is fundamental to understanding population movement patterns and supporting real-world geospatial applications such as urban planning, epidemic response, and transportation analysis. Recent works treat la...

📖 Read original article


216. TADPO: Reinforcement Learning Goes Off-road ​

Author: Zhouchonghao Wu, Raymond Song, Vedant Mundheda, Luis E. Navarro-Serment, Christof Schoenborn, Jeff Schneider
Published: 7/15/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2603.05995v2 Announce Type: replace-cross Abstract: Off-road autonomous driving poses significant challenges such as navigating unmapped, variable terrain with uncertain and diverse dynamics. Addressing these challenges requires effective long-horizon planning and adaptable control. Reinforcem...

📖 Read original article


217. Hierarchical Latent Structures in Data Generation Process Unify Mechanistic Phenomena across Scale ​

Author: Jonas Rohweder, Subhabrata Dutta, Iryna Gurevych
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2603.06592v2 Announce Type: replace-cross Abstract: Contemporary studies in mechanistic interpretability have uncovered many puzzling phenomena in the neural information processing of Transformer-based language models, such as induction heads, function vectors, and the Hydra effect. Some of th...

📖 Read original article


218. APPLV: Adaptive Planner Parameter Learning from Vision-Language-Action Model ​

Author: Yuanjie Lu, Beichen Wang, Zhengqi Wu, Yang Li, Xiaomin Lin, Chengzhi Mao, Xuesu Xiao
Published: 7/15/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2603.08862v2 Announce Type: replace-cross Abstract: Autonomous navigation in highly constrained environments remains challenging for mobile robots. Classical navigation approaches offer safety assurances but require environment-specific parameter tuning; end-to-end learning bypasses parameter ...

📖 Read original article


219. Learning When to Trust in Contextual Social Bandits ​

Author: Majid Ghasemi, Mark Crowley
Published: 7/15/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2603.13356v2 Announce Type: replace-cross Abstract: Robust reinforcement learning typically assumes that feedback sources are either globally trustworthy or corrupted within a fixed global budget. We identify a more subtle failure mode that escapes this dichotomy, which we call \emph{Contextua...

📖 Read original article


220. The TIME Machine: On The Power of Motion for Efficient Perception ​

Author: Mantas Skackauskas, Xinyue Hao, Laura Sevilla-Lara
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2605.23045v3 Announce Type: replace-cross Abstract: Video representation learning has seen tremendous progress in recent years. This has been driven by many factors, including the scale of training and the success of visual models trained contrastively with language. While these factors have p...

📖 Read original article


221. Entropy Equivalence Testing ​

Author: Cl'ement L. Canonne, Yash Pote, Jonathan Scarlett, Joy Qiping Yang
Published: 7/15/2026, 4:00:00 AM
Categories: cs.DS, cs.DM, cs.IT, cs.LG, math.IT, math.ST, stat.TH

arXiv:2605.23225v2 Announce Type: replace-cross Abstract: We introduce the problem of \emph{entropy equivalence testing} for probability distributions, a relaxation of the well-studied closeness testing problem, where the distribution testing algorithm is now only required to distinguish, given samp...

📖 Read original article


222. The Score Hamiltonian: Mapping Diffusion Models to Adiabatic Transport ​

Author: Peter Halmos, Boris Hanin
Published: 7/15/2026, 4:00:00 AM
Categories: math-ph, cs.AI, cs.LG, math.MP, physics.data-an

arXiv:2606.05217v3 Announce Type: replace-cross Abstract: We exhibit an exact correspondence between sampling with score-based diffusion models and adiabatic transport of ground states for a family of Schr"odinger operators we call Score Hamiltonians, built from the learned score's quantum potentia...

📖 Read original article


223. Does Topic Sentiment Cause Perceived Ideology? Comparing Human and LLM Annotations in Political News Articles ​

Author: Upasana Chatterjee
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2606.06715v2 Announce Type: replace-cross Abstract: We ask whether topic sentiment has a causal effect on perceived political ideology, and whether the answer depends on who assigns the ideology label. Using articles from AllSides, paired with shared sentiment annotations from Llama-3.3-70b-ve...

📖 Read original article


224. ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories ​

Author: Siyuan Luo, Nairong Zheng, Lin Zhou, Tiankuo Yao, Shengyou Yuan, Haojia Yu, Cong Pang, Jiapeng Luo, Lewei Lu
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2606.11520v4 Announce Type: replace-cross Abstract: Training capable OS agents requires data that simultaneously captures structured user intents, multi-turn task delegation, and grounded tool execution--properties absent from existing datasets. We propose ISE (Intent -> Simulate -> Execute), ...

📖 Read original article


225. Atlas H&E-TME: Scalable AI-Based Tissue Profiling at Expert Pathologist-Level Accuracy ​

Author: Kai Standvoss, Miriam H"agele, Rosemarie Krupar, Julika Ribbat-Idel, Jennifer Altsch"uler, Gerrit Erdmann, Hans Pinckaers, Evelyn Ramberger, Madleen Drinkwitz, 'Ad'am N'arai, Alexander M"ollers, Katja Lingelbach, Sebastian Kons, Lukas H"onig, Recepcan Adig"uzel, Joana Bai~ao, Alberto Megina Gonzalo, Marius Teodorescu, Marie-Lisa Eich, Paolo Chetta, Shakil Merchant, Verena Aumiller, Simon Schallenberg, Andrew Norgan, Klaus-Robert M"uller, Lukas Ruff, Maximilian Alber, Frederick Klauschen
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2606.12346v2 Announce Type: replace-cross Abstract: Hematoxylin and eosin (H&E) staining is the cornerstone of histopathology, yet scalable, quantitative analysis of H&E whole-slide images (WSIs) remains a central challenge in computational pathology. We present Atlas H&E-TME, an AI-based syst...

📖 Read original article


226. Significance-First Splitting: Aligning Treatment Heterogeneity Detection with Honest Estimation ​

Author: Pantelis Z. Hadjipantelis, Weng Man Chiang, Karthik Nagesh
Published: 7/15/2026, 4:00:00 AM
Categories: stat.ME, cs.LG, stat.ML

arXiv:2607.03999v2 Announce Type: replace-cross Abstract: Estimating heterogeneous treatment effects (CATE) requires simultaneously detecting effect modification and quantifying estimation uncertainty. Existing tree-based methods make an uneasy trade-off: significance-based approaches (Radcliffe and...

📖 Read original article


227. AgentLens: Production-Assessed Trajectory Reviews for Coding Agent Evaluation ​

Author: Andrey Podivilov, Vadim Lomshakov, Sergey Savin, Matvei Startsev, Roman Pozharskiy, Maksim Parshin, Sergey Nikolenko
Published: 7/15/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE

arXiv:2607.06624v2 Announce Type: replace-cross Abstract: We present AgentLens, a production-assessed benchmark for interactive code agents. Most code-agent benchmarks reduce a run to a single bit -- did the task pass? -- but the people who actually use these agents experience the entire trajectory:...

📖 Read original article


228. Automated Tensor Scheduling for Hybrid CPU-GPU LLM Inference on Consumer Devices ​

Author: Yangyijian Liu, Hongyi Ye, Mingyang Li, Wu-jun Li
Published: 7/15/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG

arXiv:2607.10183v2 Announce Type: replace-cross Abstract: Running large language models on consumer devices such as laptops and desktops is challenging because model weights often exceed GPU memory capacity, making offloading inference necessary to extend effective model capacity with CPU memory. Ex...

📖 Read original article


229. Toward Production-Ready Federated Learning in Healthcare: Privacy, Orchestration, and Governance in MLOps ​

Author: Sakshi Gorkhali, Jonesh Shrestha
Published: 7/15/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2607.10467v2 Announce Type: replace-cross Abstract: Healthcare organizations often cannot freely centralize patient data because medical records are sensitive, regulated, and institutionally controlled. Federated learning offers a practical alternative by allowing hospitals and clinics to trai...

📖 Read original article


230. RecRec: Recursive Refinement for Sequential Recommendation ​

Author: Pervez Shaik, Prosenjit Biswas, Abhinav Thorat, Ravi Kolla, Niranjan Pedanekar
Published: 7/15/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2607.10541v3 Announce Type: replace-cross Abstract: Sequential recommender systems typically infer user preferences through single-pass encoding of interaction histories without iterative refinement, relying on increasingly deep architectures to capture complex patterns. In this work, we revis...

📖 Read original article


231. Can a Language Model Learn Facts Continually in Its Weights? ​

Author: Charles O'Neill
Published: 7/15/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2607.11020v2 Announce Type: replace-cross Abstract: Continual learning promises a language model that keeps acquiring knowledge after training, with each new fact written into its weights. Whether weight writes can support accumulation remains undecided. We follow invented facts written into Q...

📖 Read original article


232. Bringing Back Rule Induction to Fluid Intelligence Research? An Initial Validation of the ARC-AGI Benchmark in Humans ​

Author: Jasmin Thelen, Oliver Wilhelm
Published: 7/15/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2607.11263v3 Announce Type: replace-cross Abstract: Two competing perspectives on fluid intelligence (gf) measures propose that performance is primarily constrained either by working memory capacity or by the ability to induce novel relations. The first perspective is currently dominant in mea...

📖 Read original article


233. SKooP: Symmetric Koopman Predictions for Faster and More Generalizable Legged Robot Locomotion with Reinforcement Learning ​

Author: Evelyn D'Elia, Weishu Zhan, Giulio Turrisi, Giulio Romualdi, Giuseppe L'Erario, Raffaello Camoriano, Wei Pan, Daniele Pucci
Published: 7/15/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2607.11624v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) algorithms classically suffer from poor sample efficiency. In robotics, a recent line of work has emerged addressing this problem by encoding physics priors in the learning process. However, most of these approache...

📖 Read original article


234. Imputation-free transformer learning enables robust Alzheimer's disease prediction and calibrated uncertainty quantification across heterogeneous clinical cohorts ​

Author: Christelle Schneuwly Diaz, Narmina Baghirova, Duy-Thanh Vu, Duy-Cat Can, Gilles Allali, Philippe Ryvlin, Oliver Y. Ch'en
Published: 7/15/2026, 4:00:00 AM
Categories: q-bio.NC, cs.LG

arXiv:2607.11656v2 Announce Type: replace-cross Abstract: Accurate diagnostic classification and disease-severity prediction for Alzheimer's disease are hampered by the incompleteness and heterogeneity of real-world clinical data. Left unaddressed, these barriers prevent reliable disease modelling a...

📖 Read original article