Skip to content

arXiv cs.LG - 2026-09-01 ​

593 items collected.


1. ERR+: Sequential Entropy Resolution for Efficient and Decisive LLM Reasoning ​

Author: Xin Jiang, Minhao Wang, Wen Wu, Zhentao Xie, Shangheng Du, Jinxin Shi, Jiabao Zhao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.28771v1 Announce Type: new Abstract: Large reasoning models achieve strong performance on complex tasks by generating extended chain-of-thought (CoT) traces via reinforcement learning with verifiable rewards (RLVR). While current RLVR methods have achieved strong results with correctness-...

📖 Read original article


2. Unsupervised Latent Space Alignment with Hyperspherical Geodesic Matching ​

Author: Cameron Ryan, Vivek Sivaraman Narayanaswamy, Kowshik Thopalli, Shusen Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.28840v1 Announce Type: new Abstract: Independently trained neural networks tend to encode the same data with similar latent geometries. These latent geometries are not directly compatible, yet they can be nearly the same up to some class of transformations. While there exists many methods...

📖 Read original article


3. Curvature Cryptanalysis of Smooth Transformer Feed-Forward Networks ​

Author: Munawar Hasan, Apostol Vassilev
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2608.28843v1 Announce Type: new Abstract: We show that smooth two-layer feed-forward networks (FFNs) expose an additional structural model extraction channel under a chosen-input raw-output oracle at the FFN branch; consider transformer FFN branches with GELU or SiLU activations under chosen-i...

📖 Read original article


4. Equivariant Sheaf Neural Networks: Learning Geometric Transport on Graphs ​

Author: Alessio Borgi, Mario Severino, Fabrizio Silvestri, Pietro Li`o
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.28853v1 Announce Type: new Abstract: Equivariant graph neural networks provide a principled way to model geometric systems, but efficient first-order architectures remain limited in how vector information can be transformed as it moves across a graph. We introduce \textsc{ESNN}, an Equiva...

📖 Read original article


5. The Halt Vector: Internalizing a Causal Steering Intervention for Efficient Reasoning ​

Author: Dylan Jayabahu, Tinuade Adeleke
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.28859v1 Announce Type: new Abstract: Reasoning models do not stop when they know the answer. On DeepSeek-R1-Distill-Qwen-7B the chain of thought runs about twice as long as the model's own answer probability takes to settle, and how much of that excess is removable varies from problem to ...

📖 Read original article


6. Conservative Hybrid Graph Networks for Process Systems with Learned Routing ​

Author: Paolo Guida
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.28896v1 Announce Type: new Abstract: Industrial process networks do not maintain a single effective topology while operating: streams are throttled or bypassed, and units move between idle, transition, and active regimes. Models of such systems are typically trained on measured state traj...

📖 Read original article


7. Off-Policy Evaluation for Semantic ID Recommenders: Does the Model's Own Code Hierarchy Help? ​

Author: Artem Betlei
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.28905v1 Announce Type: new Abstract: Generative recommenders increasingly emit semantic IDs (SIDs): each item is a short sequence of hierarchical discrete codes from a residual quantizer, decoded autoregressively. Before spending scarce A/B-test, a team may decide offline which decoder or...

📖 Read original article


8. Learning-Theoretic Foundation for General Coded Computing: The Straggler Setting ​

Author: Parsa Moradi, Behrooz Tahmasebi, Mohammad Ali Maddah-Ali
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.28910v1 Announce Type: new Abstract: Coded computing has emerged as a powerful paradigm for mitigating the impact of straggling workers in distributed computing systems. However, existing coded-computing schemes are predominantly designed for the exact recovery of highly structured comput...

📖 Read original article


9. SemKV: Semantic Mixed-Precision KV Cache Quantization Guided by the Quality Cliff for Long-Context LLM Inference ​

Author: Daeha Lee, Do-Hyung Kim, Jae-Hong Kim
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IT, math.IT

arXiv:2608.28911v1 Announce Type: new Abstract: The key-value (KV) cache is the dominant memory bottleneck of long-context large language model (LLM) inference, growing linearly with context length. We show that uniform KV quantization on a fractional-bit grid does not degrade gracefully: under a pr...

📖 Read original article


10. RankShift: In-Database Detection and Explanation of Categorical Shifts ​

Author: Omair Shafi Ahmed
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.28922v1 Announce Type: new Abstract: A login service can receive its usual number of failed sign-ins while one source grows from 2% to 30% of them. The same pattern appears in system logs when a rare event template becomes common while the message rate stays stable. These events change wh...

📖 Read original article


11. Revisiting the Provable-Auditable Privacy Gap of DP-SGD ​

Author: Saloni Modi, Srivi Balaji, Yusong Zhu, Gautam Kamath, Kevin Tian
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.28934v1 Announce Type: new Abstract: Differential privacy (DP) has traditionally been used to provide theoretical upper bounds on an algorithm's stability to changing its training data. In modern private machine learning applications, achieving strong tradeoffs between utility and theoret...

📖 Read original article


12. From the Loss Landscape to Diverse Feature Learning in Neural Networks ​

Author: David Aram Yunis
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.28948v1 Announce Type: new Abstract: Over the course of the last decade, neural networks have grown from an academic curiosity to moving the markets of nations. Despite this explosion in both research and deployment, relatively little is understood about how they achieve the solutions the...

📖 Read original article


13. Continuity-Free Near-Minimax Leading-Order Regret for CVaR-UCBVI ​

Author: Yuanlong Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.28960v1 Announce Type: new Abstract: For finite-horizon tabular CVaR reinforcement learning, prior work proves a $\widetilde{O}(\tau^{-1}\sqrt{SAK})$ leading regret bound for arbitrary normalized return laws and the sharper $\widetilde{O}(\sqrt{SAK/\tau})$ rate under a density lower bound...

📖 Read original article


14. V2TATC: A Joint Voice-Trajectory Embedding Framework and Dataset for Air Traffic Controller Situational Awareness ​

Author: Louis Brusset, Mathurin Petit, Jordan Kam, Alexandre Bayen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, eess.AS

arXiv:2608.28981v1 Announce Type: new Abstract: As air traffic volumes in the National Airspace System continue to expand, in particular in the low altitude airspaces, the need for scalable decision support tools used by air traffic controllers will also require more development. This article introd...

📖 Read original article


15. Effective Graph and Rank-based Contextual Embeddings for Textual and Multimedia Data ​

Author: Thiago C'esar Castilho Almeida, Gustavo Rosseto Let'icio, Lucas Pascotti Valem, Andr'e Freitas, Daniel Carlos Guimar~aes Pedronette
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.IR

arXiv:2608.29001v1 Announce Type: new Abstract: In a data-driven world, efficiently organizing and mapping relationships between objects is crucial. Graphs are powerful tools for modeling these connections, being widely used in social networks, telecommunications, and biology. However, graph-based m...

📖 Read original article


16. Context-Aware Interpretable Representations for Retrieval and Graph Convolutional Network Classification ​

Author: Thiago C'esar Castilho Almeida, Gustavo Rosseto Let'icio, Vinicius Atsushi Sato Kawai, Daniel Carlos Guimar~aes Pedronette
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CV, cs.IR

arXiv:2608.29004v1 Announce Type: new Abstract: The advances in visual information modeling and representation during the last decades are remarkable, mainly supported by Convolutional Neural Networks, Transformer-based, and Foundation Models. Despite this progress, critical challenges regarding the...

📖 Read original article


17. Hybrid Semantic Context-Enhanced Ensemble Learning for Wind Power Ramp-Event Forecasting and Uncertainty-Aware Evaluation ​

Author: Momina Liaqat Ali, Muhammad Abid, Muhammad Abdullah, Aneela Zameer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29024v1 Announce Type: new Abstract: Wind power ramp events which are sudden, large swings in turbine output over short windows are difficult to estimate, and standard models often miss them. Hybrid forecasting approach is built which augments semantic context to ramp-event forecast. Rath...

📖 Read original article


18. Flow-JEPA: Flow Matching for Robust Latent Dynamics in JEPA World Models ​

Author: Yanchen Huo, Ziying Song, Yadan Luo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29029v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) have shown strong potential for learning compact predictive representations, and LeWorldModel (LeWM) extends this paradigm to reconstruction-free latent world modeling from pixels. However, its determini...

📖 Read original article


19. NVE: A Separability and Coverage-Aware Internal Validation Metric for Biclustering ​

Author: Paritosh Tiwari, I Navin Kumar, James C. Bezdek, Punit Rathore
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29045v1 Announce Type: new Abstract: Biclustering, or co-clustering, aims to discover coherent submatrices by grouping rows and columns of a data matrix simultaneously. This local two-dimensional structure makes validation more difficult than in ordinary clustering, where internal indices...

📖 Read original article


20. Sparse Koopman Autoencoders Identify Local Dynamical Regimes in Multibasin Systems ​

Author: Aidan Li, Uday Kiran Reddy Tadipatri, Mahan Fathi, Sarath Chandar, Ross Goroshin
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, math.DS

arXiv:2608.29057v1 Announce Type: new Abstract: Koopman autoencoders (KAEs) seek a higher-dimensional latent representation in which nonlinear dynamics evolve linearly. However, many interesting systems have multiple basins of attraction, and both theoretical and empirical work has shown these multi...

📖 Read original article


21. PathBridger: Subgoal Bridges for Offline Goal-Conditioned Reinforcement Learning ​

Author: Soohyun Choi, Seonvin Cho, Songnam Hong
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2608.29061v1 Announce Type: new Abstract: Offline goal-conditioned reinforcement learning (GCRL) aims to learn policies for reaching diverse goals entirely from fixed trajectory data. Long-horizon offline GCRL remains challenging because sparse goal-reaching signals must be propagated over man...

📖 Read original article


22. Selective Disclosure of Hidden Directives in Reasoning Models: Behavioral Asymmetry and Steering ​

Author: Zimo Shi, Xander Tifft, Wen Xing
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29070v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning traces are increasingly proposed as a mechanism for AI oversight: a monitor inspecting a model's reasoning can, in principle, detect misbehavior invisible from outputs alone. This assumes CoT surfaces what a model is in...

📖 Read original article


23. Titans-QFWP: A Regime-Aware Hybrid Quantum Fast Weight Programmer for Portfolio Optimization ​

Author: Ming-Kai Hung, Jun-Hao Chen, Yun-Cheng Tsai, Samuel Yen-Chi Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29093v1 Announce Type: new Abstract: We propose Titans-QFWP, a hybrid reinforcement learning architecture integrating a Quantum Fast Weight Programmer with Titans-style memory (Persistence, Surprise, and Forgetting) for adaptive portfolio optimization. To address high-dimensional market f...

📖 Read original article


24. Development of an Autonomous AI Coding Agent using Monte Carlo Tree Search (MCTS) and Gemini LLM Frameworks ​

Author: Pravin Game, Vipin Ramakrishnan, Prathamesh Wagh
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29096v1 Announce Type: new Abstract: The ongoing changes in software engineering requirements have created a substantial need for automated tools which can create secure source code from natural language input. The performance of traditional Large Language Models (LLMs) becomes limited by...

📖 Read original article


25. Temperature-Adaptive Transformed Teacher Matching ​

Author: Hiroaki Aizawa, Yoshikazu Hayashi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.29099v1 Announce Type: new Abstract: Temperature scaling is a core component of knowledge distillation, yet its role and effect are still not fully understood. Transformed Teacher Matching (TTM) clarifies the role of temperature scaling by applying it only to the teacher distribution and ...

📖 Read original article


26. PathGuide: Dynamic Classifier-Free Guidance via On-Policy Transport Alignment ​

Author: Avishag Nevo, Tamir Hazan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.29107v1 Announce Type: new Abstract: While modern generative models excel at modeling complex data, precise inference-time control in conditional generation remains a critical challenge. Classifier-free guidance (CFG) is a primary mechanism for such control, yet it is typically treated as...

📖 Read original article


27. Explainable Machine Learning for Broadband Adoption Disparities: Tract-Level Prediction and SHAP-Based Factor Profiling ​

Author: Xiao Han
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29110v1 Announce Type: new Abstract: The United States has allocated approximately $65 billion through the Infrastructure Investment and Jobs Act for broadband expansion, yet evidence-based methods for targeting these investments remain underdeveloped. This paper presents an explainable m...

📖 Read original article


28. Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space ​

Author: Qiancheng Zhou, Ruizhe Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.29188v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) substantially improves single-sample accuracy (pass@1) but causes the policy's solution space to contract, diminishing the returns of test-time scaling. In this work, we investigate where inside a r...

📖 Read original article


29. HalluPrism: When Multimodal Uncertainty Should Diagnose, Not Decide ​

Author: Aman Prakash, Sourish Dasgupta, Tanmoy Chakraborty
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29193v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) can assign similar confidence to answers that fail for different reasons. We propose HalluPrism, a behavioral diagnostic that re-runs an answer after visual degradation, blank-image replacement, and grounding or...

📖 Read original article


30. PokaiTrainer: Scaling Belief-State Search to Competitive Pok\'emon VGC ​

Author: Max Yu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GT

arXiv:2608.29197v1 Announce Type: new Abstract: Decision-time equilibrium search carried poker to superhuman play, but it has so far relied on tractable subgames: a handful of actions per decision, chance confined to card deals, one player moving at a time. Competitive Pok'emon in its official doub...

📖 Read original article


31. RL-FAT: Reinforcement Learning for Fair Adversarial Training ​

Author: Tejaswini Medi, Levan Mikeladze, Margret Keuper
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29247v1 Announce Type: new Abstract: Deep neural networks remain highly vulnerable to adversarial perturbations, and adversarial training (AT) has become a widely used approach for improving robustness. However, improvements in average robust accuracy often mask substantial class-wise dis...

📖 Read original article


32. Adaptive Multi-Branching for Shallow Decision Tree Induction ​

Author: Hanul Park, Jeonghoon Choi, Juseong Kim, Sanghun Sel, Giltae Song
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29262v1 Announce Type: new Abstract: Decision trees are attractive for tabular prediction tasks because each prediction follows an interpretable sequence of feature-threshold tests. Under a strict maximum-depth budget, however, conventional binary trees can be under-expressive, since each...

📖 Read original article


33. When Do Larger Batches Help Scale LLM Reinforcement Learning? ​

Author: Ziniu Li, Jinbo Wang, Guanhua Huang, Feiyuan Zhang, Pengbo Li, Alex Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29296v1 Announce Type: new Abstract: Larger batches reduce the variance of stochastic gradients per update and are therefore often expected to accelerate training. Yet whether this statistical benefit translates into lower wall-clock time-to-target remains unclear, because each update con...

📖 Read original article


34. A Spectral Identifiability Threshold for Dissipative Rate Recovery from Truncated Liouvillian Spectra ​

Author: Yujun Ji, Somyajit Chakraborty
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, physics.comp-ph, quant-ph

arXiv:2608.29302v1 Announce Type: new Abstract: Open quantum systems lose energy and phase coherence through different dissipative processes, but these processes can produce overlapping dynamical signatures. The Liouvillian spectrum summarizes how such a system relaxes, yet it is not obvious how muc...

📖 Read original article


35. MEL: Coordinate-Preserving EEG Tokenization for fMRI Translation ​

Author: Xiangyu Liu, Zeting Yan, Zhitong Yin, Boyang Li, Xi Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29304v1 Announce Type: new Abstract: Translating electroencephalography (EEG) into functional magnetic resonance imaging (fMRI) is important for medical neuroimaging, clinical brain-state monitoring, and multimodal neural decoding, because it aims to infer spatially organized hemodynamic ...

📖 Read original article


36. Information-Based Calibration of Uncertainty Quantification in Product-of-Experts Gaussian Process Models ​

Author: Yean Hoon Ong, Paolo Barucca, Wei Pan, Jun Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29349v1 Announce Type: new Abstract: Gaussian process (GP) regression with a single global GP (GP-glo) incurs cubic computational cost, limiting scalability to large datasets. Product-of-experts GP models (GP-pro), which combine local GP models to capture global correlations, alleviate th...

📖 Read original article


37. Spatial Entropy based Partitioning for Spatiotemporal Graph Unlearning ​

Author: Qiming Guo, Wenbo Sun, Ye Wang, Wenlu Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29360v1 Announce Type: new Abstract: Spatiotemporal graphs underpin applications such as traffic forecasting, weather forecasting, and healthcare monitoring. Privacy regulations such as the GDPR and the CCPA require the complete removal of unauthorized data from trained models, but achiev...

📖 Read original article


38. Unlearning on Spatio-Temporal Graphs through Subgraph Virtual Edge Reconstruction ​

Author: Qiming Guo, Wenbo Sun, Chen Pan, Ye Wang, Wenlu Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CR

arXiv:2608.29369v1 Announce Type: new Abstract: Spatio-temporal graphs are widely used in modeling complex dynamic processes such as temporal forecasting, molecular dynamics, and healthcare monitoring. Recently, stringent privacy regulations such as GDPR and CCPA have introduced significant new chal...

📖 Read original article


39. Fully Distributed GNE Algorithms for Multi-Robot Placement without Consensus on Multipliers ​

Author: Shao-An Yin, Mingyi Hong, Nicola Elia
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GT, cs.MA, cs.RO

arXiv:2608.29388v1 Announce Type: new Abstract: Recent machine learning research has increasingly focused on equilibrium analysis in non-cooperative games rather than solely on optimal solutions. Many such problems involve shared constraints and can be formulated as Generalized Nash Equilibrium Prob...

📖 Read original article


40. Where Induction Runs Out: Description-Length Difficulty and the Memorisation Gap in Integer-Sequence Benchmarks ​

Author: Sabilashan Ganeshan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.SC

arXiv:2608.29411v1 Announce Type: new Abstract: Integer sequences from the On-Line Encyclopedia of Integer Sequences (OEIS) are increasingly used to benchmark mathematical reasoning in language models. We ask what such benchmarks actually measure, using an exactly computable reference learner: two-p...

📖 Read original article


41. Scalable Clinical Data Infrastructure and Comparative ML Evaluation for Hospitalisation Risk Prediction in Elderly Patients with Multiple Long-Term Conditions using CPRD ​

Author: Asra Aslam, Volodymyr Chapman, Maurice M. O'Connell, Aseel S. Abuzour, Michael Abaho, Danushka Bollegala, Gary Leeming, Eduard Shantsila, Andrew Clegg, Lauren E. Walker, Iain Edward Buchan, Samuel D. Relton
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29419v1 Announce Type: new Abstract: Deep learning architectures are increasingly proposed for patient trajectory modeling in electronic health records (EHRs), yet their advantage over simpler, more interpretable models is rarely subjected to rigorous empirical scrutiny in real-world clin...

📖 Read original article


42. One Capability or Many? Testing the Economic Validity of Frontier AI Evaluation ​

Author: Louis Yiven Zhu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CY, cs.SE

arXiv:2608.29420v1 Announce Type: new Abstract: Frontier-model leaderboards now rank systems based on economic benchmarks, tests of how well models carry out professional tasks from software engineering to banking workflows, and those rankings inform what organisations buy, what regulators scrutinis...

📖 Read original article


43. Behavioral Latency as Weak Event-Time Supervision for EEG Reaction-Time Decoding ​

Author: Anuar Aimoldin, Ayana Mussabayeva, Yedige Mussabayev, Xue Liu, Kun Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29428v1 Announce Type: new Abstract: Single-trial EEG analyses are often organized around events and latencies, yet EEG-based reaction-time (RT) prediction is posed as scalar regression on a fixed stimulus-locked window. RT is treated as a window-level label rather than timing evidence ab...

📖 Read original article


44. Does Latent Planning Survive Point Clouds? Action-Conditioned JEPA World Models for Geometric Observations ​

Author: Fabio F. Oberweger, Michael Schwingshackl
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2608.29434v1 Announce Type: new Abstract: JEPA world models make latent-space planning a practical route to control, but they are built almost exclusively on images. Whether latent prediction survives geometric observations is unclear: point clouds are sparse, unordered, and self-occluded, and...

📖 Read original article


45. SS-ESOAP: Self-Scaled Adaptive Preconditioning for Physics-Informed Learning ​

Author: Guangyuan Wang, Mads Toftrup, Sebastian Loeschcke, Yixuan Wang, Anima Anandkumar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC, stat.ML

arXiv:2608.29448v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) often face ill-conditioned objectives that limit high-accuracy training. Dense quasi-Newton methods improve local conditioning but require expensive optimizer state, while Kronecker-factored methods such as SOAP...

📖 Read original article


46. Reference-Grafting Matches Fine-Tuning at Eliciting Sandbagged Capabilities ​

Author: Linh Le, Hong Kiat Tan, David Williams-King
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29458v1 Announce Type: new Abstract: Sandbagging, in which a model deliberately underperforms on an evaluation despite retaining the underlying capability, threatens the safety evaluations that frontier-model governance depends on. The Elicitation Game found that fine-tuning elicits hidde...

📖 Read original article


47. A Causal Model for Locating and Unlocking Sandbagging in Model Organisms ​

Author: Hong Kiat Tan, Linh Le, David Williams-King
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29461v1 Announce Type: new Abstract: Sandbagging models strategically underperform on evaluations while retaining the capabilities being measured. The evaluations that guide frontier-model deployment and governance then understate what these models can do. To understand the mechanism, we ...

📖 Read original article


48. Knowledge Distillation under Teacher Misspecification: An Order-Parameter Analysis of the Gap between Teacher Mimicry and Task Performance ​

Author: Kazuyuki Hara, Hideitsu Hino
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29472v1 Announce Type: new Abstract: Knowledge distillation trains a small student model to reproduce the outputs of a large teacher model, and its progress is typically monitored through the teacher--student discrepancy. The quantity of ultimate interest, however, is the student's error ...

📖 Read original article


49. Learning Human Health and Diseases from 24-hour Wrist Movement ​

Author: Yong Wang, Dylan McGagh, Katya Broomberg, Zizheng Zhang, Jonathan Carter, Junayed Naushad, Laura Brocklebank, Yang Sun, George Nicholson, Dianjianyi Sun, Canqing Yu, Jun Lv, Maxim Barnard, Hubert Lam, Andrew Steptoe, David W. Eyre, Liming Li, Zhengming Chen, Naomi Wray, Spiros Denaxas, Gary S. Collins, Huaidong Du, Aiden Doherty, Hang Yuan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29494v1 Announce Type: new Abstract: Much of human health and function unfolds beyond the clinic, through the movements of everyday life. Wrist-worn accelerometers capture these movements continuously, yet their rich signals are often reduced to a small set of predefined behavioural summa...

📖 Read original article


50. Target-Aware State-Adaptive $p$-Dirichlet Graph Neural Regression for Non-Invasive Body-Composition Estimation ​

Author: Nadejda Drenska, Matthew Lemoine, Gowri Priya Sunkara, Yu Wang, Sri Lakshmi Sravani Devarakonda, Steven B. Heymsfield
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29496v1 Announce Type: new Abstract: Accurate estimation of body-composition outcomes, including body fat percentage (BFP), bone mineral density (BMD), and appendicular lean mass (ALM), is important for evaluating metabolic, skeletal, and muscular health. Direct assessment using dual-ener...

📖 Read original article


51. Adversarial Online Classification with a Preview ​

Author: Roi Livni, Sahil Singla
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.DS

arXiv:2608.29503v1 Announce Type: new Abstract: Worst-case online classification is governed by sequential complexity, such as Littlestone dimension, and can be impossible even for statistically simple classes, such as thresholds of VC dimension one. We study a preview model in which an oblivious ad...

📖 Read original article


52. Denoising as Projection: Constrained Optimization with Gradient-Guided Diffusion ​

Author: Runyu Zhang, Jiawei Zhang, Gioele Zardini, Saurabh Amin, Asuman Ozdaglar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC

arXiv:2608.29507v1 Announce Type: new Abstract: Diffusion models are increasingly used not only for sampling from learned data distributions, but also for generating samples that optimize task-specific objectives. A common approach is to guide the reverse diffusion process using gradients of an exte...

📖 Read original article


53. On the Plasticity Collapse in Continual Machine Unlearning ​

Author: Yingdan Shi, Xiang Xu, Kaize Ding, Alfred O. Hero, Ren Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29513v1 Announce Type: new Abstract: Machine unlearning enables deep neural networks to selectively remove the influence of specific data in response to privacy and regulatory requirements. While prior work largely studies single-shot unlearning, real-world systems must accommodate contin...

📖 Read original article


54. MedCache: Efficient and Temporally Valid Memory for Longitudinal Clinical Agents ​

Author: Hei Ting (Una), Chan, Chenwei Wu, Xueshen Liu, Boyuan Zheng, Liyue Shen, Jiasi Chen, Z. Morley Mao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.DC, cs.MA

arXiv:2608.29528v1 Announce Type: new Abstract: Longitudinal clinical agents must maintain an evolving patient state from evidence distributed across visits, time points, and specialties. However, how agent memory should be designed for this setting remains unclear. We introduce a benchmark of multi...

📖 Read original article


55. BEACON: Behavioral and Semantic Enrichment of AlphaEarth Embeddings through Tri-Modal Contrastive Learning ​

Author: Hao Tian, Heng Cai, Yifan Yang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29553v1 Announce Type: new Abstract: Geospatial foundation models such as the AlphaEarth Foundation produce compact and globally consistent representations of the Earth's surface that transfer effectively to a wide range of downstream tasks. However, because these models are trained prima...

📖 Read original article


56. Which LLM for Which Work? Budgeted Model Allocation under Uncertain Evaluation ​

Author: Hamed Khosravi, Xiaoming Huo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, math.OC, stat.ML

arXiv:2608.29560v1 Announce Type: new Abstract: A company with a fixed artificial intelligence (AI) budget must decide which large language model (LLM) handles each recurring workload. What it lacks is the quality table, how well each model performs on each workload. Given that table, the decision i...

📖 Read original article


57. Asynchronous Cooperative Online Learning for Multi-Robot Control under Computational Delays ​

Author: Xiaobing Dai, Zewen Yang, Wei Ren, Sandra Hirche
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.MA, cs.RO, cs.SY, eess.SY

arXiv:2608.29562v1 Announce Type: new Abstract: Ensuring the safe operation of multi-agent systems (MASs) under uncertain environments is crucial for cooperative robotic, where external disturbances and inaccurate dynamic models can significantly compromise performance and reliability. To address th...

📖 Read original article


58. HoopMind: A Real-Time Neural Game-Tree System for Opponent-Aware Possession Planning ​

Author: Yibo Gong, Cong Guo, Jiacheng Ding
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.HC

arXiv:2608.29563v1 Announce Type: new Abstract: School coaches prepare for opponents with game film and intuition. The analytics tools of professional teams stay out of reach. We ask how far public data can close this gap. Professional basketball is our case study, chosen for its data rather than th...

📖 Read original article


59. Event-triggered Control and Online Learning for Networked Systems under Computational Delays ​

Author: Xiaobing Dai, Armin Lederer, Zewen Yang, Sihua Zhang, Lu Wan, Yang Tang, Sandra Hirche
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.RO, cs.SY, eess.SY

arXiv:2608.29576v1 Announce Type: new Abstract: Online learning-based control is a promising approach to control uncertain systems, where unknown components are identified during operation to improve control performance. However, resource-intensive online learning algorithms introduce non-negligible...

📖 Read original article


60. Predicting the Unpredictable: LLM-powered Long-term Chaotic Time Series Forecasting under Short-term Observations ​

Author: Yuhang Yao, Bohan Jiang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29579v1 Announce Type: new Abstract: Chaotic time series forecasting is a challenging task due to its sensitivity to initial conditions and long-term unpredictability. Traditional methods typically rely on sufficient temporal trajectories to learn long-term dynamics, which limits their ap...

📖 Read original article


61. On the Resilience of Text-to-Video Diffusion Models to Hardware Faults ​

Author: Zachary Coalson, A M Aahad, Stella Doehring, Zane Ma, Sanghyun Hong
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29598v1 Announce Type: new Abstract: We present the first systematic study of the resilience of text-to-video (T2V) diffusion models under random hardware-level faults. While T2V models are widely used for automated video generation due to their ability to produce high-quality, temporally...

📖 Read original article


62. Adaptive Doubly Robust Off-Policy Evaluation for Ranking Policies under Diverse User Behavior ​

Author: Kosuke Iguchi, Ren Kishimoto
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2608.29600v1 Announce Type: new Abstract: Off-policy evaluation (OPE) of ranking policies is challenging be- cause selecting and ordering multiple items from a candidate set makes the number of possible rankings grow combinatorially with the number of candidates and the ranking length. Consequ...

📖 Read original article


63. Wide Learning: Learning to Reach Evidence ​

Author: Junzhou Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29608v1 Announce Type: new Abstract: Machine learning is usually evaluated after an evidence interface has been fixed. A dataset, sensor suite, query language, action set, or experimental protocol determines which observations can be obtained, and learning is judged by what it extracts fr...

📖 Read original article


64. Unsupervised Multi-Scale Gromov-Wasserstein Hypergraph Alignment ​

Author: Lutz Oettershagen, Honglian Wang, Aristides Gionis
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.SI

arXiv:2608.29635v1 Announce Type: new Abstract: We study unsupervised hypergraph alignment, where the goal is to infer node correspondences between two hypergraphs using only structural information, without node features, labels, seed matches, or side information. Direct higher-order formulations ca...

📖 Read original article


65. LLMODE: Aligning ODEs with LLMs via Gated Token Injection for Irregular Spatio-Temporal Forecasting ​

Author: Di Zhang, Jingyang Zhang, Ziqian Wang, Chi Zhang, Yikun Ban, Ziwei Zhang, Ruijie Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29640v1 Announce Type: new Abstract: Large language models (LLMs) have shown promise for spatio-temporal forecasting, but existing approaches often rely on regularly sampled token sequences and struggle with irregular observations because of temporal asynchrony, representation-space misal...

📖 Read original article


66. Reward-guided Fine-Tuning of One-Step Generative Models via Wasserstein Gradient Flow ​

Author: Hoseong Hwang, Woorim Han, Joungin Chun, Jinseong Park, Jaewoong Choi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29647v1 Announce Type: new Abstract: To mitigate the time complexity of generative models, one-step generative models have recently emerged through direct mapping from noise to data in a single forward pass. However, the reward-guided fine-tuning method of one-step generative models remai...

📖 Read original article


67. A Target-Centric Survey of Quantization-Aware Training ​

Author: Jiamin Song, Mengjie Zhao, Zijing Wang, Yongkang Liu, Qian Li, Shi Feng, Feiliang Ren, Daling Wang, Hinrich Sch"utze
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29667v1 Announce Type: new Abstract: The rapid development of LLMs incurs prohibitive memory footprints and intensive computational demands. Quantization-Aware Training (QAT) techniques have emerged as a promising solution to address these challenges by explicitly simulating quantization ...

📖 Read original article


68. Creation begins with understanding: LLMs as strategy designers for privacy-preserving tabular data synthesis ​

Author: Jinmeng Li, Quan Zhang, Hangting Ye, He Zhao, Firas Laakom, Dandan Guo, J"urgen Schmidhuber
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29674v1 Announce Type: new Abstract: Sharing tabular data in high-stakes domains is constrained by privacy regulations. Synthetic data offer a promising alternative, but deep generative models are costly to train and difficult to audit, while LLM-based methods often serialize records as t...

📖 Read original article


69. Last Step Matters: Early Uncertainty Cannot Predict Failure in Long-Horizon Agents ​

Author: Zongyue Li, Chengyue Yu, Lei Zang, Chenyi Zhuang, Linjian Mo, Leilei Gan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29685v1 Announce Type: new Abstract: Early failure prediction is important for long-horizon agents, as it enables timely intervention and can reduce inference and tool-use costs. Uncertainty quantification, such as verbal confidence and perplexity, offers a promising approach to detecting...

📖 Read original article


70. Higher-Dimensional Rotary Position Embedding ​

Author: Yixing Li, Ruobing Xie, Yudong Zhang, Yushi Bai, Samm Sun, Yu Cheng
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.29715v1 Announce Type: new Abstract: Transformers rely on position embedding mechanisms in long context modeling in most cases. Rotary Position Embedding (RoPE) embeds positional information with independent 2D rotations, forming relative position terms in self-attention. However, its pai...

📖 Read original article


71. GraM-Diff: A Unified Graph-Mamba Diffusion Framework for EEG-Based Alzheimer's Disease Data Generation and Diagnosis ​

Author: M. Tanveer, Ayush Singh Rana, Sanskriti Jain, Arnav Kumar, Aryaman Tiwari, A. Rahaman, A. Quadir, M. Sajid
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29755v1 Announce Type: new Abstract: Electroencephalography (EEG) is a promising, non-invasive, and cost-effective modality for Alzheimer's disease (AD) detection, but deep learning methods are limited by small and imbalanced clinical datasets. Generative augmentation offers a solution, y...

📖 Read original article


72. ECA-BLS: An Efficient Complex-Augmented Broad Learning System ​

Author: A. Rahaman, A. Quadir, M. Sajid, M. Akhtar, M. Tanveer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29763v1 Announce Type: new Abstract: Broad Learning System (BLS) is an efficient alternative to deep architectures due to its fast training, analytical learning, and strong generalization under limited data. However, existing BLS variants are confined to real-valued representations, restr...

📖 Read original article


73. PruneShift: A Framework for Evaluating Decision Reliability in Structured Pruning ​

Author: Hao Ye, Gaopeng Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2608.29765v1 Announce Type: new Abstract: Structured pruning uses surrogate objectives because direct task evaluation over every feasible mask is too expensive. Most evaluations report average surrogate error or rank correlation on broadly sampled masks. These summaries do not directly test th...

📖 Read original article


74. Structure Aware Neural Architecture Search for Mixture of Experts ​

Author: Petr Babkin, Oleg Bakhteev
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29817v1 Announce Type: new Abstract: Neural Architecture Search (NAS) has so far rarely been applied to Mixture-of-Experts (MoE) models, and existing MoE designs leave the alignment between experts and the structure of the data to emerge on its own. We propose an architecture search frame...

📖 Read original article


75. Designing for the Next Click: Bandits for Real-Time Page Layout ​

Author: Bhavtosh Rath, Harshith Narasimhamurthy, Bob Eisinger, Cole Stiegler, Adnan Awow, Amit Pande
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29850v1 Announce Type: new Abstract: E-commerce platforms increasingly personalize user experiences through machine learning, yet page layout decisions remain dominated by static rules and manual curation. We present a scalable bandit-based system that optimizes product page layouts in re...

📖 Read original article


76. Uncertainty-Driven Replay Memory for Reinforcement Learning ​

Author: Sheeraja Rajakrishnan, Alexander G. Ororbia, Travis Desell, Daniel E. Krutz
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29860v1 Announce Type: new Abstract: Uncertainty estimation provides promising capabilities for reinforcement learning (RL) agents. Notably, estimating uncertainty can reduce the training time and enable agents to obtain greater rewards over time by exploiting information related to wheth...

📖 Read original article


77. Partially Linear Autoencoders for Manifold Learning and Dimensionality Reduction ​

Author: Louen Pottier, Louis Lesueur, Anders Thorin
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29867v1 Announce Type: new Abstract: Autoencoders are widely used for nonlinear dimensionality reduction and manifold learning. While most common implementations rely on both nonlinear encoders and decoders, we investigate the specific role of the encoder and the extent to which it can be...

📖 Read original article


78. Towards an Expressivity-Normalized Energy-Demand Comparison of ANNs and SNNs ​

Author: Miriam Kranzlm"uller, Pascal Esser, Gitta Kutyniok
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29869v1 Announce Type: new Abstract: Spiking neural networks (SNNs) are often regarded as energy-efficient alternatives to artificial neural networks (ANNs), yet their advantage depends critically on both network architecture and data properties. We develop an analytical framework to comp...

📖 Read original article


79. Structural Hierarchy and Geometry in Molecular Representation Learning ​

Author: David Sulu, Lorenzo Di Fruscia, Jana M. Weber
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM

arXiv:2608.29886v1 Announce Type: new Abstract: Molecular self-supervised learning uses chemical structures to guide which molecular embeddings should be similar. We study whether explicitly encoding a molecule's Bemis-Murcko scaffold and using it to supervise the molecular embedding changes what th...

📖 Read original article


80. Sensitivity-Constrained Neural Operators for Data-Efficient Forward and Inverse Modeling of Partial Differential Equation Systems ​

Author: Abdolmehdi Behroozi, Chaopeng Shen, Daniel Kifer, Kathryn Lawson
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29888v1 Announce Type: new Abstract: Neural operators provide fast surrogates for partial differential equation (PDE) solvers, but their reliability can degrade for high-dimensional spatial inputs and inverse or repeated inference. State-only training constrains solution values but not th...

📖 Read original article


81. Joint Spatiotemporal Spectral Neural Operators for Learning PDEs on Irregular Domains ​

Author: Abdolmehdi Behroozi, Chaopeng Shen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29892v1 Announce Type: new Abstract: Learning solution operators for partial differential equations (PDEs) on irregular and geometry-dependent domains remains a central challenge in scientific machine learning. While spectral methods provide strong inductive biases for modeling global int...

📖 Read original article


82. INTERVenE: Temporal-Abstraction-Interval Based Transformers for Short-Horizon Medical Event Prediction ​

Author: Shahar Oded, Yuval Shahar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.29901v1 Announce Type: new Abstract: Electronic Health Record (EHR) prediction models in the intensive care unit must learn from sparse and irregular measurements while preserving the clinical meaning of time and supporting transparent decision-making. We present INTERVenE, a family of Tr...

📖 Read original article


83. Diffusion-Based Inverse Design of Dielectric Resonator Metasurfaces for Shaping Smart Electromagnetic Environments ​

Author: M. Tsukerman, K. Grotov, D. Vovchuk, P. Ginzburg
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, physics.app-ph, physics.comp-ph

arXiv:2608.29907v1 Announce Type: new Abstract: Future wireless systems are expected to transform the surrounding space from a passive propagation medium into a smart electromagnetic environment, where engineered surfaces control wave propagation, support wireless sensing, and create programmable el...

📖 Read original article


84. On the Recoverability of Private Information Unlearning in Large Language Models ​

Author: Shicheng Hu, Runzhi Tian, Ziqiao Wang, Yongyi Mao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.29943v1 Announce Type: new Abstract: Large language models (LLMs) can memorize sensitive information, raising serious privacy concerns. Machine unlearning offers a potential solution to remove such information, but it remains unclear whether existing methods truly erase it or merely hide ...

📖 Read original article


85. Robust Broad Learning System with Wave Loss for Classification under Data Uncertainty ​

Author: Mushir Akhtar, A. Varshney, A. Quadir, A. Rahaman, M. Tanveer, Mohd. Arshad
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29983v1 Announce Type: new Abstract: Broad Learning System (BLS) offers an efficient alternative to deep architectures by enabling fast learning through randomized feature mapping and closed-form solutions. However, its reliance on squared error loss makes it highly sensitive to noise, ou...

📖 Read original article


86. The Intervention Gap in Latent World Models ​

Author: Donna Vakalis
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.29998v1 Announce Type: new Abstract: Planning-time intervention fidelity is a distinct, measurable property of a learned world model: whether the model's own open-loop transitions move task variables the way matched environment interventions do. In the settings we test, it is neither reve...

📖 Read original article


87. Error Detection for PET/CT Radiology Reports: Domain-Specific vs Large Language Models ​

Author: Hermione Warr, Harry Anthony, Lilli J Freischem, Yasin Ibrahim, Daniel R McGowan, Konstantinos Kamnitsas
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30021v1 Announce Type: new Abstract: Errors in radiology reports can adversely affect patient treatment, yet automated report quality assurance remains challenging because errors are often subtle and require domain expertise to detect. Although large language models (LLMs) have recently b...

📖 Read original article


88. Multiclass Linear Perceptrons with Multiplicative Margins ​

Author: Dmitri Rachkovskij, Evgeny Osipov, Olexander Volkov, Daswin De Silva, Denis Kleyko
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.NE

arXiv:2608.30028v1 Announce Type: new Abstract: This paper introduces a family of multiclass linear Perceptron classifiers with a multiplicative margin mechanism (MMPerc), as an alternative to standard margin-free and additive margin Perceptrons. The multiplicative formulation enforces classificatio...

📖 Read original article


89. Forget or Fine-tune? A Comparative Study of Machine Unlearning Strategies for Noisy Label Correction ​

Author: Jo~ao L. P. Santana, Filipe R. Cordeiro
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30046v1 Announce Type: new Abstract: Noisy labels remain a critical challenge for training deep neural networks, since memorizing incorrect labels degrades generalization. Once noisy samples are identified after training, the standard solution is to retrain the model from scratch on the c...

📖 Read original article


90. When 3D Gaussian Splatting Recovers Real Surfaces ​

Author: Songhe Wang, David Johnathan Miller
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.30054v1 Announce Type: new Abstract: When does 3D Gaussian Splatting (3DGS) recover the true scene surface rather than just overfitting view-dependent appearance? We answer this by developing a mathematical framework based on a first-hit rendering abstraction that cleanly isolates geometr...

📖 Read original article


91. How do World Models and Policies Compose in LLM Agents? A Joint Spectral and Behavioral Account ​

Author: Ruize Xu, Xiao Yu, Yujin Tang, Chenming Shang, Nikhil Singh
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.30067v1 Announce Type: new Abstract: How do LLM agents come to both understand environments they act in and master tasks set within them? Through controlled experiments combining world-model training (next-state prediction) and policy training (reward maximization), we investigate this qu...

📖 Read original article


92. Selection, Representation, and Execution in Sparse Fourier Neural Operators ​

Author: Abdul Qadir Ibrahim, Martin Burger
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2608.30070v1 Announce Type: new Abstract: Sparse representations are often expected to make models smaller and also reduce inference cost. For Fourier Neural Operators (FNOs), these objectives are not equivalent or do not always align: removing parts of the learned operator can leave the under...

📖 Read original article


93. Tracing Generated Samples to Training-Data Clusters in Flow-Matching Models ​

Author: Rania Briq, Ohad Fried, Michael Kamp, Stefan Kesselheim
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.30081v1 Announce Type: new Abstract: Understanding which training samples influence a generated image is an important problem in generative modeling. In flow matching, training samples influence the generated image through the velocity field along the generation trajectory. Removing sampl...

📖 Read original article


94. A Lightweight Phenology-Aware YOLOv5 Framework for Tomato Growth Stage Detection in Resource-Constrained Bhutanese Greenhouse Environments ​

Author: Sherab Gocha, Sou Nobukawa
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30088v1 Announce Type: new Abstract: Accurate detection of tomato growth stages is essential for stage-specific greenhouse management and precision agriculture. In Bhutan, greenhouse cultivation is affected by altitude variability, large diurnal temperature fluctuations, diffuse illuminat...

📖 Read original article


95. SMOTE-VAR: An Uncertainty-Aware Oversampling Method for Predicting Depression Remission in University Students ​

Author: Dang Nguyen, Arun Kumar A V, Taylor A. Braund, Wu Yi Zheng, Debopriyo Bal, Leonard Hoon, Jill Newby, Helen Christensen, Svetha Venkatesh, Alexis Whitton, Sunil Gupta
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30102v1 Announce Type: new Abstract: University students experience disproportionately high rates of common mental health conditions, such as depression, which can impair learning, social functioning, and overall well-being. Although lifestyle interventions such as mindfulness and physica...

📖 Read original article


96. Graph4BiLO: Graph Neural Network Approximation for Bilevel Mixed-Integer Linear Optimization ​

Author: Jessica D. Elrefaei, Kaixun Hua, Seungbae Kim, Hoang Nam Tran, Juan S. Borrero
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30103v1 Announce Type: new Abstract: Bilevel mixed-integer linear optimization problems model hierarchical decision processes in which a leader anticipates the optimal response of a follower. Although expressive, these problems are computationally challenging because lower-level optimalit...

📖 Read original article


97. Supraglacial Lake Fate Is Knowable Long Before the Season Ends ​

Author: Emam Hossain, Md Osman Gani
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30113v1 Announce Type: new Abstract: A supraglacial lake on the Greenland Ice Sheet ends its melt season in one of four ways: it drains rapidly through a hydrofracture, drains slowly across the surface, refreezes in place, or is buried by late-season snowfall. Which one occurs decides whe...

📖 Read original article


98. TPR-Attention for Combinatorial Generalization ​

Author: Melisa Civeleko\u{g}lu, Isabeau Pr'emont-Schwarz
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE

arXiv:2608.30124v1 Announce Type: new Abstract: Systematic generalization remains a significant challenge in deep learning. In particular, combinatorial generalization - generalizing to new configurations of known factors of variation - is effortless for humans but difficult for standard neural arch...

📖 Read original article


99. Converse and Collision-Based Achievability for Node Localization with Hybrid Distance-Spectral Graph Positional Encodings ​

Author: Zimo Yan, Yifan Li, Hao Li, Zheng Xie, Chang Liu, Zheming Tu, Yuan Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30152v1 Announce Type: new Abstract: Graph positional encodings are widely used in graph neural networks and graph Transformers, yet it remains unclear when the code itself can identify nodes. We study a hybrid distance-spectral encoding that combines anchor-distance profiles with quantiz...

📖 Read original article


100. Reinforcement Learning for Symbolic Equation Solving ​

Author: Kevin P O Keeffe
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30162v1 Announce Type: new Abstract: We present a reinforcement-learning agent that solves symbolic equations step by step, covering both nonlinear closed equations (radicals, exponentials, trigonometric) and a controlled class of restricted-open families requiring a change of variables (...

📖 Read original article


101. Benchmarking Peptide-Protein Affinity Prediction Across Peptide and Target Shifts ​

Author: Jiaxin Tian, Darren An, Jun Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, q-bio.QM

arXiv:2608.30175v1 Announce Type: new Abstract: Peptide-protein affinity models are often evaluated with a single data split, obscuring whether they interpolate among measurements for observed targets or generalize across peptide or target shifts. We integrated three sources of quantitative peptide-...

📖 Read original article


102. Certified Safety Radii in Forecast-Error Space for Wasserstein Distributionally Robust Small Signal Stability-Constrained AC Optimal Power Flow via Lifted Spectrahedral Containment ​

Author: Ziqi Zhang, Xi Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30201v1 Announce Type: new Abstract: Directly robustifying small-signal stability in AC optimal power flow is challenging since the stability boundary in the original uncertainty space is implicit, highly nonconvex, and changes with the operating decision. This paper exploits an alternati...

📖 Read original article


103. Diffusion-Based Refinement for Kilometer-Scale Probabilistic Precipitation Nowcasting ​

Author: Dohyun Park, Changhoon Song, Tengyuan Chang, Yoo-Geun Ham, Youngjoon Hong
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, physics.ao-ph

arXiv:2608.30205v1 Announce Type: new Abstract: Localized extreme precipitation is a major trigger of urban flash floods and landslides, yet producing nowcasts that combine fine spatial detail with probabilistic uncertainty remains challenging. Here we introduce exPreCast-ENS, a conditional residual...

📖 Read original article


104. Strong Drafts Need Compact Memories: Long-Context Speculative Decoding with Compressed KV Cache ​

Author: Tong Yuan, Chengxi Liao, Zeyi Wen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30252v1 Announce Type: new Abstract: Long-context LLM applications such as document summarization and multi-turn agents require generation from prefixes spanning tens of thousands of tokens, making decoding latency a major bottleneck. Speculative decoding (SD) reduces latency without chan...

📖 Read original article


105. Exact Recovery Thresholds for Weighted Data Selection in Vector-Valued Linear Regression ​

Author: Guangjian Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.TH

arXiv:2608.30254v1 Announce Type: new Abstract: We resolve the threshold part of Question 4 of the COLT 2025 open problem "Data Selection for Regression Tasks" of Hanneke, Moran, Shlimovich and Yehudayoff. In vector-valued linear regression with square loss $\ell_{(x,y)}(W)=|Wx-y|_2^2$, where $x\in...

📖 Read original article


106. Multivariate Scientific Data Compression with Learned Cross-Variable Latent Decorrelation and Autoregressive Entropy Modeling ​

Author: Liangji Zhu, Anand Rangarajan, Sanjay Ranka
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30262v1 Announce Type: new Abstract: Scientific simulations generate collections of physical fields with heterogeneous statistics and dependencies, yet learned compressors often encode those fields independently or rely on a shared encoder without explicitly modeling the structure that re...

📖 Read original article


107. BCPPO: Bachelier-Inspired Constrained Proximal Policy Optimization for Tail-Risk-Aware Safe Reinforcement Learning ​

Author: Dongsheng Hou, Yanqiao Chen, Yuhan Rui
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30283v1 Announce Type: new Abstract: Expected-cost constraints can still permit rare, high-cost events. Monte Carlo conditional value at risk (CVaR) gradients can be noisy at high confidence, whereas critics that model an outcome distribution add complexity. We propose BCPPO (Bachelier-In...

📖 Read original article


108. CateKV: On Sequential Consistency for Long-Context LLM Inference Acceleration ​

Author: Haoyun Jiang, Haolin Li, Jianwei Zhang, Fei Huang, Qiang Hu, Minmin Sun, Shuai Xiao, Yong Li, Junyang Lin, Jiangchao Yao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30295v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated strong capabilities in handling long-context tasks, but processing such long contexts remains challenging due to the substantial memory requirements and inference latency. In this work, we discover that ce...

📖 Read original article


109. Tail-Replay: Escaping the Curse of Linear Attention in Prefix Caching for Hybrid LLMs ​

Author: Yirui Liu, Ruoling Qi, Xuaner Wu, Penghang Liu, Jian Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30310v1 Announce Type: new Abstract: Hybrid large language models interleave full-attention layers with linear-attention layers to reduce the cost of long-context inference. This structure complicates prefix caching: full-attention key-value caches are token-addressable, whereas linear-at...

📖 Read original article


110. Context Staircase: Signature-Aligned Dynamics of Token Embeddings under Small Initialization ​

Author: Junjie Yao, Liangkai Hang, Zhi-Qin John Xu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.30315v1 Announce Type: new Abstract: Token embeddings are the basic representational units that connect discrete tokens with continuous computation in language models. Although modern language models learn embeddings from random initialization through gradient-based training, the dynamica...

📖 Read original article


Author: Donggyu Min, Dong-Kyu Kim
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30317v1 Announce Type: new Abstract: Online dynamic origin-destination (OD) matrix estimation (DODE) calibrates time-dependent OD demand to reproduce observed link-flow trajectories. In online, OD demand should be estimated from current observations and propagated network states while sub...

📖 Read original article


112. Generative multi-domain transfer learning for fault detection in data-scarce wind turbines ​

Author: Stefan Jonas, Angela Meyer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30323v1 Announce Type: new Abstract: Normal behavior models have shown promise for reliable fault detection in wind turbines. However, these unsupervised anomaly detection models require sufficient fault-free training data to learn the normal operation behavior of turbines. Under data sca...

📖 Read original article


113. Learning PDE Time-Stepping with Neural Cellular Automata ​

Author: Esha Saha, Hao Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.30328v1 Announce Type: new Abstract: Classical numerical solvers for partial differential equations (PDEs) are computationally expensive to solve repeatedly across varying initial conditions, motivating the need for learned surrogates. In this paper, we propose a trainable Neural Cellular...

📖 Read original article


114. Coarse composition suffices: tabular in-context learning for multi-activity antimicrobial peptide profiling ​

Author: Raunak Kumar, Anuj Pal, Dhruvi Solanki, Parikshit Pareek, Juhi Singh, Jitin Singla
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, q-bio.BM

arXiv:2608.30337v1 Announce Type: new Abstract: Antimicrobial peptides (AMPs) often act against multiple pathogen classes, making multi-label activity prediction a more realistic screening target than binary antimicrobial classification. The ESCAPE benchmark formalizes this setting, but leading appr...

📖 Read original article


115. Beyond Churn: Predicting Financial Fragmentation in Retail Banking with Temporal Machine Learning ​

Author: Ananyaa Chopra, Brandon Xu, Brendan Yuen, Lauren Zung, Sarabroop Aulakh
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30364v1 Announce Type: new Abstract: Retail banking attrition is usually represented as a terminal binary event, even though client relationships often weaken earlier through partial movements of deposits, investments, and recurring activity to external financial institutions. This paper ...

📖 Read original article


116. Mode Connectivity Beyond Classifiers: Evidence from Generative and Contrastive Models ​

Author: Chengzheyi Yao, Yongzhao Zhang, Yongding Tian
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30366v1 Announce Type: new Abstract: The loss landscape of Deep Neural Networks (DNNs) exhibits highly complex and non-convex properties. Recent studies have revealed the phenomenon of mode connectivity, demonstrating that independently trained network modes can be connected via a continu...

📖 Read original article


117. Beat-Synchronous Tokenization for ECG Transformers ​

Author: Ahmed Sameh, Nolan Wilson, Max Enderlein, Yogatheesan Varatharajah
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2608.30367v1 Announce Type: new Abstract: Transformer-based electrocardiogram (ECG) models commonly tokenize waveforms into fixed temporal patches. Though convenient, fixed patching can split heartbeat structures across token boundaries. We study beat-synchronous tokenization as a physiologica...

📖 Read original article


118. Convergence rates for the RMSprop optimizer with full control of the hyperparameters ​

Author: Steffen Dereich, Arnulf Jentzen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, math.OC, math.PR

arXiv:2608.30382v1 Announce Type: new Abstract: Popular adaptive stochastic gradient descent (SGD) methods to train artificial intelligence (AI) systems include the RMSprop, the Adam, and the AdamW optimizers, where the adaptivity parts in Adam and AdamW basically just coincide with RMSprop. Such ad...

📖 Read original article


Author: Rastislav Lenhardt, Teodora Dobos, Thomas Vecchiato, Jiri Isa, Igor Ginzburg
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2608.30384v1 Announce Type: new Abstract: By introducing RSLM (Rotated Scaled Lloyd-Max), a family of training-free vector quantization codecs compressing embeddings to 1--4 bits per dimension, we reduce memory cost and memory bandwidth of a typical large-scale Approximate Nearest Neighbor (AN...

📖 Read original article


120. DASC: Decay-Aware State Compression for Hybrid Linear-Attention Serving ​

Author: Yanqi Yu, Pingwei Sun, Jianchao Tan, Tao Zhang, Yuchen Xie, Xunliang Cai, Yao Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30386v1 Announce Type: new Abstract: Hybrid linear-attention architectures have recently scaled to large open-weight models, offering quality competitive with full attention while substantially reducing key/value (KV) cache growth. However, their in-place recurrent-state updates complicat...

📖 Read original article


121. Uncertainty of Vision Medical Foundation Models ​

Author: Haoxu Huang, Narges Razavian
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30390v1 Announce Type: new Abstract: Accurate uncertainty estimation is essential for machine learning systems de- ployed in high-stakes domains such as medicine. Traditional approaches primarily rely on probability outputs from trained models (point predictions), which provide no formal ...

📖 Read original article


122. Foundation Models Meet Agriculture: Challenges Beyond Pretraining ​

Author: Vishal Nedungadi, Xingguo Xiong, Marc Ru{\ss}wurm, Ioannis N. Athanasiadis
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30392v1 Announce Type: new Abstract: Global food security and sustainable climate action increasingly rely on robust, scalable agricultural monitoring. Earth observation foundation models have emerged as powerful, label-efficient tools across general remote sensing domains, yet early atte...

📖 Read original article


123. TopGQ: Fast GNN Post-Training Quantization Leveraging Topology Information ​

Author: Dain Kwon, Kanghyun Choi, Hyeyoon Lee, Sunjong Park, Seoyong Lee, Sukjin Kim, Jinho Lee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30394v1 Announce Type: new Abstract: Existing GNN quantization methods suffer from considerable quantization overhead, which severely limits their practical usage in real-world scenarios. To this end, we present TopGQ, an accurate post-training GNN quantization framework, alleviating redu...

📖 Read original article


124. Locally-Guided Actor-Critic: Training a Goal-conditioned Actor with a Subgoal-aware Critic ​

Author: Olivier Serris, St'ephane Doncieux, Olivier Sigaud
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30406v1 Announce Type: new Abstract: Goal-conditioned reinforcement learning struggles with long horizons when rewards are sparse. While a planner can provide subgoals to guide a low-level policy, its use at test time may introduce practical subgoal management difficulties. An alternative...

📖 Read original article


125. No Equivariant Architecture Covers All Equivariant Attention ​

Author: T=ikun ^Ong
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, math.AG, math.RT

arXiv:2608.30417v1 Announce Type: new Abstract: We give a complete characterization of equivariant multi-head self-attention (MHSA): if an MHSA layer is equivariant to a symmetry group $G$, then $G$ can only act by permuting head-clusters, with QK and OV matrices satisfying an equivariance constrain...

📖 Read original article


126. Confounding Masquerading as Improvement: A Systematic Evaluation of Offline Reinforcement Learning for Stroke Antithrombotic Treatment in a 129,000-Patient Registry ​

Author: Kihun Rhee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.30442v1 Announce Type: new Abstract: Recent offline reinforcement learning (RL) studies report policies that outperform physician decisions on clinical outcomes. We conduct a systematic, partially crossed evaluation of five offline RL algorithm families and 14 reward designs in 44,894 pos...

📖 Read original article


127. PRIME: Mitigating Subgroup Optimization Competition in Shared CTR Top Networks with Plug-in Residual Input-Conditioned Mixture of Expert ​

Author: Heng Yao, Siyun Hou, Tianying Liu, Yulou Shu, Yong He, Chuan Yuan, Kaibin Qiu, Guowei Chen, Jiayu Zhao, Chao Yu, Ke Ding
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.IR

arXiv:2608.30449v1 Announce Type: new Abstract: Click-through rate (CTR) models vary in feature-interaction design, yet their top networks usually remain a single multilayer perceptron shared by all examples. Heterogeneous user, item, and context subgroups therefore update the same parameters; weakl...

📖 Read original article


128. Self-Supervised Pretext Tasks for Infant Cry Analysis: A Controlled Comparison and a Cautionary Result on Donateacry ​

Author: Luigi Simeone
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30456v1 Announce Type: new Abstract: We compare six self-supervised pretext tasks for infant cry analysis under a fixed budget, meaning the same compact encoder of 1.17M parameters, the same 115 hours of license-verified public pretraining audio, and the same evaluation protocol for every...

📖 Read original article


129. Learning Where Outcomes Change:Credit-Addressable Reasoning for Multimodal Geometry ​

Author: Jiani Guo, Junjie Wang, Jie Wu, Pengxiang Zhao, Dongdong Zhang, Shaohan Huang, Yujiu Yang, Furu Wei
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.30457v1 Announce Type: new Abstract: Multimodal geometry reasoning requires VLMs to extract precise visual relations and preserve them through multi-step deduction. Existing free-form traces obscure the decisions that determine the answer, and trajectory-level reinforcement learning distr...

📖 Read original article


130. ToxLens: A Reproducible Graph-Learning Framework for Leakage-Aware, Uncertainty-Calibrated Molecular Toxicity Prediction ​

Author: Magnus H. Str{\o}mme, Alex G. C. de S'a, David B. Ascher
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30472v1 Announce Type: new Abstract: Molecular toxicity prediction is increasingly used to prioritise compounds before experimental testing, but conventional benchmark performance can overstate practical utility when structurally related molecules occur across training and test folds. We ...

📖 Read original article


131. Measuring Memory and Generalization as Separable Geometric Channels: The Topo^2 Framework ​

Author: Zhanbo Zhang, Ming Liu, Qing Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30487v1 Announce Type: new Abstract: Deep networks trained on noisy labels simultaneously generalize on clean data and memorize flipped labels. These are usually conflated as pressures on one capacity. We present Topo^2, a measurement framework that makes them causally separable, measurab...

📖 Read original article


132. When the Martingale Never Stops Firing: Anytime-Valid Gating on Real Forecast Streams ​

Author: Weijia Han, Lisha Qu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ME, stat.ML

arXiv:2608.30502v1 Announce Type: new Abstract: Machine learning systems are increasingly corrected while they run, and the decision of when to intervene is increasingly delegated to statistical monitors. Anytime-valid inference promises evidence that can be acted on at any moment, exactly the guara...

📖 Read original article


133. Tensor Methods for Language Models: From Token Representation to Training, Adaptation, Inference, Compression, and Interpretability ​

Author: Matvei Tarasov, Salman Ahmadi-Asl, Andre L. F. de Almeida, Andrzej Cichocki
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30505v1 Announce Type: new Abstract: Large language models (LLMs) are built from structured high-dimensional objects such as token representations, weights, adaptation updates, caches, and activations, whose multilinear structure is underexploited by the conventional matrix-centric view. ...

📖 Read original article


134. Trajectory-Initialized Neural Double Q-Routing for Large-Scale Overhead Hoist Transport Systems ​

Author: Cheng Gu, Qiusheng Zhao, Anbang Liu, Shaochong Lin, Max Z. J. Shen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, math.OC

arXiv:2608.30512v1 Announce Type: new Abstract: Large-scale industrial robot fleets share constrained physical infrastructure, making vehicle travel times dependent on safety separation, intersection access, downstream blocking, and station contention. We study this problem in overhead hoist transpo...

📖 Read original article


135. PAC: Progress-Augmented Advantage Curriculum for Multi-Task Reinforcement Learning of LLMs ​

Author: Yuanqiang Yu, Yanzhao Zheng, Zhentao Zhang, Tianze Xu, Chao Ma, Jihuai Zhu, Jiashun Liu, Xinle Deng, Baohua Dong, Hangcheng Zhu, Ruohui Huang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30528v1 Announce Type: new Abstract: Reinforcement learning (RL) is used to improve the reasoning abilities of LLMs, while training data span heterogeneous tasks. However, most RL post-training pipelines rely on fixed or manually designed task mixtures, even though task usefulness changes...

📖 Read original article


136. Q-Strata: Hierarchical Bit Allocation for Mixed-Precision Quantization of Mixture-of-Experts LLMs ​

Author: Deokjae Lee, Sihun Chu, Hyun Oh Song
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30564v1 Announce Type: new Abstract: Mixed-precision quantization (MPQ) assigns a different bitwidth to each linear layer of a large language model (LLM) to minimize the quantization-induced quality loss under a fixed budget, but Mixture-of-Experts (MoE) models contain these layers in eve...

📖 Read original article


137. Collapsibility of Performance Metrics in Clinical Predictive AI ​

Author: Jo~ao Matos, Ben Van Calster, Richard D. Riley, Paula Dhiman, Gary S. Collins
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30568v1 Announce Type: new Abstract: Background: Population level assessments of predictive artificial intelligence (AI) can conceal performance disparities across subgroups. Fairness evaluations commonly rely on performance analyses across subgroups. However, some performance metrics are...

📖 Read original article


138. The Safety Relay in Roleplay Jailbreaks: A Component-Resolved Causal Analysis of Harm Recognition and Refusal ​

Author: Md Mokarram Chowdhury, Ernie Chang, Yang Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30585v1 Announce Type: new Abstract: Large language models are trained to follow instructions while refusing harmful requests. Jailbreaks exploit this balance to elicit content a model would ordinarily reject. Roleplay jailbreaks are especially concerning: the harmful request can remain v...

📖 Read original article


139. State of Health Estimation using Convolutional and Bidirectional LSTM Neural Networks tuned by Bayesian Optimization ​

Author: Panagiotis Eleftheriadis, Foivos Georgios Kyrgios, Sonia Leva
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30593v1 Announce Type: new Abstract: In this research, a novel framework is proposed for the SOH estimation, which employs a hybrid deep learning architecture of a concatenation of a Convolution Neural Network (CNN) and a Bidirectional Long Short-Term Memory (BiLSTM) Neural Network (NN) w...

📖 Read original article


140. PLC-DPO: Posterior Label Correction in Noisy and Ambiguous Preference Optimization ​

Author: Boryeong Cho, Sumyeong Ahn, Se-Young Yun
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.30597v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) simplifies alignment through pairwise comparisons but assumes all observed preferences are reliable. Real data often violates this assumption, leading to reversed, weak, or ambiguous labels that cause harmful policy...

📖 Read original article


141. MolLedger: An Additive Graph Neural Network with Chemically Grounded ADME Attributions ​

Author: Christina X. Ji
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30636v1 Announce Type: new Abstract: Optimizing absorption, distribution, metabolism, and excretion (ADME) is an important part of small molecule drug discovery. Many machine learning models have been built to predict ADME properties to facilitate this optimization process, but explaining...

📖 Read original article


142. Three Steps at a Time: Learning Representations from Action Sequences in Contrastive RL ​

Author: Michal Korniak, Kamil Dybek, Benjamin Eysenbach, Marco Bagatella, Micha{\l} Bortkiewicz
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30640v1 Announce Type: new Abstract: While self-supervised approaches to reinforcement learning have achieved strong results by learning representations of states and actions, a key open question is the time scale over which actions should be modeled. Departing from the standard formulati...

📖 Read original article


143. Season-Aware Hybrid Convolutional-Transformer for Antarctic Sea Ice Concentration Forecasting ​

Author: Danyang Li, John Taylor, Thang Bui, Quanling Deng
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30654v1 Announce Type: new Abstract: Antarctic sea ice concentration (SIC) forecasting is an important yet challenging task due to the coexistence of complex spatial structure, long-range temporal dependencies, and strong seasonal variability. Conventional convolution-based models are eff...

📖 Read original article


144. CoMPASS: Collaborative Molecular Property Prediction via Adaptive Small-Large Model Synergy ​

Author: Wentao Li, Jiangjie Qiu, Yijun Li, Leyi Zhao, Xiaonan Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30674v1 Announce Type: new Abstract: Accurate molecular property prediction requires both statistical reliability and chemical reasoning. Graph neural networks can be calibrated directly on labeled assays but remain limited by the coverage of their training data. Large language models (LL...

📖 Read original article


145. Learning Materials Properties from Scarce Labels and Unlabeled Crystals ​

Author: Wentao Li, Yizhe Chen, Jiangjie Qiu, Yijun Li, Leyi Zhao, Xiaonan Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30682v1 Announce Type: new Abstract: Learning materials properties from scarce labels and unlabeled crystals is a central challenge for data-driven materials discovery. We present SemiMat, a controlled benchmark for semi-supervised materials property regression, and MatRank, a reliability...

📖 Read original article


146. Liquid Gated Attention ​

Author: Yiheng Jiang, Yuanbo Xu, Yongjian Yang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30695v1 Announce Type: new Abstract: Real-world time series often exhibit irregular sampling and extended temporal horizons, requiring models to capture continuous-time dynamics across arbitrary intervals without prohibitive scaling costs. Discrete-time methods collapse variable time inte...

📖 Read original article


147. Learning Dynamics of Logits Debiasing for Long-Tailed Semi-Supervised Learning ​

Author: Yue Cheng, Jiajun Zhang, Xiaohui Gao, Weiwei Xing, Zhanxing Zhu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30699v1 Announce Type: new Abstract: Long-tailed distributions are prevalent in real-world semi-supervised learning (SSL), where pseudo-labels tend to favor majority classes, leading to degraded generalization. While many long-tailed semi-supervised learning (LTSSL) methods have been prop...

📖 Read original article


148. Kolmogorov--Arnold against bounded translations ​

Author: Sviatoslav V. Dzhenzher
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, math.FA

arXiv:2608.30710v1 Announce Type: new Abstract: Historically originating from Hilbert's 13th problem, the Kolmogorov-Arnold representation theorem (KART) has recently experienced a major revitalisation through its applications to neural networks, specifically Kolmogorov-Arnold Networks (KANs). While...

📖 Read original article


149. Tracing distinguishability through transformer processing with stochastic LayerNorm ​

Author: Kieran Murphy
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30720v1 Announce Type: new Abstract: Representational similarity is foundational to analyses of deep networks, yet distances between point-valued representations are not intrinsically tied to downstream function: nearby states may produce different behaviors, while distant states may beha...

📖 Read original article


150. BAITBENCH: Measuring Agent Reward Hacking with Optional Shortcuts Planted in ML Tasks ​

Author: Pradyumna Shyama Prasad, Meiri Anto, Leon Eshuijs, Julian Moncarz, Kaustubh Kislay, Juan J. Vazquez
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30724v1 Announce Type: new Abstract: LLM agents are increasingly used to run autonomous ML experiments, iterating on target metrics with little human oversight. Prior work has documented reward hacking in these environments, bringing into question the validity of produced research and the...

📖 Read original article


151. E-Commerce Bench: Evaluating LLM Agents on Long-Horizon Autonomous Business Operation ​

Author: Wei Fan, Xinjie Shen, Xudong Guo, Jianhong Tu, Yang Su, Yinger Zhang, Lianghao Deng, Fengyu Wang, Baohua Dong, Yangqiu Song, Dayiheng Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.30730v1 Announce Type: new Abstract: Long-horizon agentic tasks go beyond chaining short tasks over more interaction turns. Their evolving dynamic environments and long-range dependencies require Large Language Models (LLMs) to continually explore, learn from experience, and adapt their p...

📖 Read original article


152. Functional Degeneracy in Neural Networks: Measurement and Pruning ​

Author: Maria Matveev, Pascal Esser, Ayush Bharadwaj, Lucius Bushnaq, Gitta Kutyniok
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30741v1 Announce Type: new Abstract: A central question in modern machine learning is how much a trained model can be compressed without changing its behavior, to reduce the memory, compute and energy required to deploy it. To study this, we quantify functional degeneracy through the beha...

📖 Read original article


153. TDDM-Melatt: A Decoupled Memory and Diffusion Framework for Generalizable Encrypted Traffic Classification ​

Author: Ze Chen, Qiming Yu, Zijia Song, Guozheng Yang, Wei Yan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30745v1 Announce Type: new Abstract: The widespread adoption of encrypted traffic poses severe challenges to current security situational awareness systems based on network traffic monitoring. In existing dataset-driven training and testing studies, limitations such as shortcut learning i...

📖 Read original article


154. Do VLMs Share Safety Neurons Across Modalities? ​

Author: Jiaxuan Li, Jiahao Zhang, Duc Minh Vo, Huy H. Nguyen, Pride Kavumba, Koki Wataoka
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30750v1 Announce Type: new Abstract: Vision-language models (VLMs) can comply with harmful requests delivered through images, even when their LLM backbones would refuse the same content in text. While prior work characterizes these jailbreaks empirically or at the representation level, ho...

📖 Read original article


155. PRACTICE: From Experience to Expertise in Self-Evolving Embodied Agents ​

Author: Ziyi Bai, Siqi Li, Tinglei Huang, B"orje F. Karlsson
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30760v1 Announce Type: new Abstract: Recent studies have shown that multimodal large language models (MLLMs) can serve as embodied agents, translating language instructions and visual observations into executable plans. However, building agents that can continually improve through interac...

📖 Read original article


156. T3S: Improving Multi-Task Reinforcement Learning with Task-Specific Feature Selector and Scheduler ​

Author: Yuanqiang Yu, Tianpei Yang, Yongliang Lv, Yan Zheng, Jianye Hao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30765v1 Announce Type: new Abstract: Multi-task reinforcement learning (MTRL) is a technique to train multiple tasks simultaneously, where previous works usually train a single model to solve different tasks by sharing parameters across various tasks. However, these methods are faced with...

📖 Read original article


157. TrainSDC: Characterizing and Mitigating Silent Data Corruption in Large Language Model Training ​

Author: Zhipeng Xia, Haotian Xu, Siyu Yun, Liqi Lin, Hu Liu, Yu Li, Cheng Zhuo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30769v1 Announce Type: new Abstract: LLM training is increasingly vulnerable to silent data corruption (SDC), yet existing protection methods largely treat Transformer computations uniformly because their vulnerability remains poorly understood. We present the first systematic characteriz...

📖 Read original article


158. Reciprocity Separates Gradient Flow from Rotation in Conservative Physical Learning ​

Author: Ruiwu Niu, Xiaowen Bi, Micha"el Antonie van Wyk
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, nlin.AO

arXiv:2608.30778v1 Announce Type: new Abstract: Physical learning lets a trainable material or network use its own physical response to carry error signals, reducing the need for a separately programmed backward computation. We ask what determines whether such a system follows conventional gradient ...

📖 Read original article


159. Geometric Attractor Monitoring: A Robust and Frugal Framework for Multi-modal Industrial Robotic Cycles ​

Author: Martin Bonsergent-Brachet, Jesse Read, Dany Abboud
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30804v1 Announce Type: new Abstract: Monitoring the health of heterogeneous industrial robot fleets is severely challenged by the multi-modal nature of their operational cycles and a persistent scarcity of run-to-failure data. Standard data-driven approaches, particularly deep learning ar...

📖 Read original article


160. What Emerges and What Breaks in Self-Play Driving ​

Author: Laur Sisask, Ardi Tampuu, Tambet Matiisen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30819v1 Announce Type: new Abstract: Training autonomous driving policies through pure self-play has recently shown promising results. Following Gigaflow and Puffer- Drive, we train driving policies in a similar self-play fashion, but extend the models from MLPs to Transformers and train ...

📖 Read original article


161. Deploying DeepSeek 175B Locally on a Single Consumer-Grade RTX 4060 Laptop with 32GB RAM for 200k-Scale Protein-Ligand Virtual Screening ​

Author: Rui Xiao, Yili Xu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30877v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have demonstrated exceptional performance in protein-ligand interaction prediction, but state-of-the-art pipelines for large-scale virtual screening almost exclusively rely on high-end GPU clusters with h...

📖 Read original article


162. Fine-Tuning Low-Bit Models with Gradient in Quantized Code Space ​

Author: Shiguang Wu, Zhouchen Lin, Quanming Yao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30908v1 Announce Type: new Abstract: Fine-tuning Low-bit models aims to adapt a quantized model while keeping the final deployed checkpoint in the same low-bit form. This setting is practically important as it reduces memory and inference cost for storage and deployment. Under this constr...

📖 Read original article


163. S3C-LLM: Skill-Code Guided Agentic Language Models for Spectrum-to-Structure Elucidation ​

Author: Xuanle Zhao, Xinyuan Cai, Xiang Cheng, Bo Xu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.30910v1 Announce Type: new Abstract: Spectroscopic structure elucidation is central to molecular analysis, but recent Large Language Model (LLM)-based methods mostly formulate it as direct spectrum-to-SMILES generation. Although this paradigm can leverage paired spectral data, it does not...

📖 Read original article


164. Selection-Aware Stress Testing for Interactive Agents ​

Author: Yang Xu, Chenang Li, Jiefu Zhang, Haixiang Sun, Zhou Li, Vaneet Aggarwal
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.AP, stat.ML

arXiv:2608.30916v1 Announce Type: new Abstract: Agent evaluations often use one benchmark to choose a workflow and then search for task types where its advantage weakens, so both conclusions are selected from the same data. We introduce Selection-Aware Semantic Stress Testing (\SASST{}), which learn...

📖 Read original article


165. Towards Stream Learning on Embedded Systems: Benchmarking the Memory Consumption of Stream Learning Methods ​

Author: Sebastian Buschj"ager, Nuwan Gunasekara, Heitor Murilo Gomes
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.PF

arXiv:2608.30923v1 Announce Type: new Abstract: Stream learning is commonly evaluated through predictive performance and adaptation to concept drift. However, sustained operation of a stream learner also requires predictable and bounded resource usage even on long streams. This requirement becomes e...

📖 Read original article


166. Nonparametric Contextual Pricing and Inventory Learning under Censored Demand ​

Author: Zean Han, Jing Liang, Ruihan Lin, Zezhen Ding, Jiheng Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30944v1 Announce Type: new Abstract: In online retailing, when a product sells out, a retailer often sees only the units sold, not how many customers would have bought it had inventory been available. However, the inventory level determines how much demand is revealed, and this informatio...

📖 Read original article


167. Reproducible macroscopic dynamics in a closed-loop human-AI learning system ​

Author: Minlin Wu (Tianli Qiming AI Research Institute, Sichuan Qiming Daren Technology Co., Ltd., Chengdu, China), Xu Fang (Tianli Qiming AI Research Institute, Sichuan Qiming Daren Technology Co., Ltd., Chengdu, China), Yicheng Zhang (Swiss AI Laboratories, Blonay, Switzerland), Chenyu Zhou (Tianli Qiming AI Research Institute, Sichuan Qiming Daren Technology Co., Ltd., Chengdu, China), Zhiyi Liu (Tianli Qiming AI Research Institute, Sichuan Qiming Daren Technology Co., Ltd., Chengdu, China)
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, nlin.AO

arXiv:2608.30946v1 Announce Type: new Abstract: Closed-loop human-AI systems generate high-dimensional behavioural trajectories whose collective dynamics remain obscure. Using 297,915 learners' adaptive-tutoring histories, we define semantic order variables before model fitting and test them in user...

📖 Read original article


168. One Policy Is Enough: Single-Agent Reinforcement Learning Outperforms Tree Search for Chemistry Tool Learning ​

Author: Armin Dariani, Sifan Wu, Bang Liu, Entao Yang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.30952v1 Announce Type: new Abstract: Chemistry questions often demand exact computation and database lookups that a language model cannot supply from its parameters, so it must reach for external tools. Tool use here is a three-part problem: select the right tool from a large pool, fill i...

📖 Read original article


169. Singular Curvature in ReLU Training:Differentiation and the Gradient-Flow Limit Need Not Commute ​

Author: Xiaoyang Li, Runni Zhou
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30960v1 Announce Type: new Abstract: Gradient descent (GD) is explicit Euler for gradient flow, but a state-accurate continuous-time surrogate need not remain accurate after differentiation. At every fixed nonresonant step size, ordinary automatic differentiation exactly differentiates th...

📖 Read original article


170. A Universal Context-Reuse Layer for Cross-Model KV Sharing ​

Author: Yi Li, Dongming Jiang, Yi Zhao, Bingzhe Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.30963v1 Announce Type: new Abstract: Modern large language model (LLM) serving systems increasingly operate over repeated or shared context, yet each model typically performs its own prefill computation even when another model has already processed the same input. Existing KV-cache reuse ...

📖 Read original article


171. A Human-in-the-Loop Autonomous Agent for Industry Time Series Forecasting ​

Author: Xiaoyu Tao, Mingyue Cheng, Ze Guo, Bokai Pan, Qi Liu, Shijin Wang, Enhong Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30976v1 Announce Type: new Abstract: Real-world time-series forecasting is rarely a one-shot model invocation: practitioners must formulate tasks, connect data and models, incorporate domain expertise, assess prediction plausibility, and communicate uncertainty. Specialized forecasting mo...

📖 Read original article


172. Sparse Competition during Training For the Emergence of Specialized Modules ​

Author: Baptiste Rossigneux, Karim Haroun
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.30978v1 Announce Type: new Abstract: Modularity in deep neural networks has been proposed as a means of improving both interpretability and training by promoting disentangled representations and reducing redundancy. In this work, we study the emergence of modular structure through competi...

📖 Read original article


173. Controlling Refusal Behavior of LLMs via Stiefel-Constrained Rotation Steering ​

Author: Kirill Bunin, Dmitry Bylinkin, Vladimir Aletov, Daniil Medyakov, Vladimir Solodkin, Aleksandr Beznosikov
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.30986v1 Announce Type: new Abstract: Activation steering has emerged as a lightweight approach for controlling model refusal at inference time. A growing line of research explores trainable rotations of activations to develop geometrically principled intervention mechanisms. However, exis...

📖 Read original article


174. Language-Informed Flow Matching for Trend-Guided Structure-Based 3D Molecular Generation ​

Author: Tianyu Gao, Zhikai Su, Jiashu Li, Wenjun Gao, Zichuan Ying, Zhe Zhao, Fei Zhang, Ye Wei
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.31009v1 Announce Type: new Abstract: Structure-based drug design (SBDD) requires ligands that satisfy both 3D target affinity and 1D chemical validity. Existing controllable generation methods often rely on task-specific fine-tuning or externally imposed sampling-time guidance, adding cos...

📖 Read original article


175. TSPFN: A Temporal Tabular Foundation Model for Physiological Time Series Classification ​

Author: J'er'emie Stym-Popper, Cl'ement Rambour, Federica Granese, Nicolas Thome, Olivier Bernard
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2608.31013v1 Announce Type: new Abstract: Designing models that generalize effectively in low- to medium-data regimes remains a primary challenge in medical machine learning, particularly for physiological time-series classification. While tabular foundation models such as TabPFN offer an attr...

📖 Read original article


176. Normalized Low-Rank Adaptation ​

Author: Jiale Kang, Ziyin Yue, Zheng Zhan, Yangyi Huang, Weiyang Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.31036v1 Announce Type: new Abstract: While low-rank adaptation (LoRA) is widely used for parameter-efficient model adaptation, how to regularize its training dynamics for stable and effective optimization remains underexplored. Because LoRA initializes the up-projection to zero, its early...

📖 Read original article


177. Rotational Equivariance in Machine Learning: A Comprehensive Tutorial ​

Author: Peter Lippmann, Fred A. Hamprecht
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.31045v1 Announce Type: new Abstract: Rotational symmetry is one of the most important structural principles in machine learning on 3D data. In applications ranging from physics and materials science to 3D computer vision, predictions should not depend on an arbitrary choice of coordinate ...

📖 Read original article


178. Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement ​

Author: Yi Ding, Ruqi Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.31046v1 Announce Type: new Abstract: On-policy distillation (OPD) offers dense token-level supervision as an alternative to the sparse outcome-level advantages of reinforcement learning with verifiable rewards (RLVR). However, the teacher scores student-generated trajectories that are inh...

📖 Read original article


179. Universal Transformers for Circuit Computations: Perfect Length Generalization in Tiny Transformers ​

Author: Takuya Ito, Ruchir Puri, Murray Campbell, Parikshit Ram
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.31067v1 Announce Type: new Abstract: Learning generalizable algorithmic computations remains a challenge for neural networks, as reflected in persistent failures on compositional and length generalization benchmarks. We present a provably correct, transformer parameterization (with only 2...

📖 Read original article


180. A Model with No Head and Many Thoughts ​

Author: Nikita Koriagin, Yaroslav Aksenov, George Bredis, Gleb Gerasimov, Nikita Balagansky, Daniil Gavrilov
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2608.31069v1 Announce Type: new Abstract: Large language models decode by projecting hidden states through a large vocabulary head at every step. This operation is computationally costly and forces all reasoning to be expressed in discrete tokens. We introduce Soft Latent Thinking, a method th...

📖 Read original article


181. Sycophantic Agreement Transfers with Neutral Data via Contrastive Preference Optimization ​

Author: Camila Blank, Zhuofan Ying, Christopher Potts, Peter Hase, Jing Huang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.31079v1 Announce Type: new Abstract: Sycophantic agreement refers to a behavior in which language models excessively affirm the user, often at the cost of factual accuracy. Although sycophantic agreement is a well-known failure of model alignment, there is limited understanding of how it ...

📖 Read original article


182. Stress-Testing Efficient Responsible-AI Evaluation: When Compute Savings Change Benchmark Conclusions ​

Author: Ahmed El Kady, Aravind Narayanan, Rehana Noorani, Yani Ioannou, Shaina Raza
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.31108v1 Announce Type: new Abstract: Efficient evaluation changes the protocol used to support claims about model behavior, yet it is rarely tested whether those claims remain stable after the evaluation itself is made cheaper. We stress-test conclusion robustness in responsible-AI benchm...

📖 Read original article


183. On the Complexity of the Compatibility Problem for Succinctly Encoded Conditional Distributions ​

Author: Guy Emerson
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CC, math.PR

arXiv:2608.31120v1 Announce Type: new Abstract: The motivation for this paper is the investigation of the trade-offs implicit in probabilistic models used in machine learning. Models are often used to make predictions in the form of conditional probabilities. However, a pair of conditional distribut...

📖 Read original article


184. Sharp Approximation Rates for Neural Networks with Affine Latent Parameterizations ​

Author: Shijun Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2608.31157v1 Announce Type: new Abstract: Many parameter-efficient methods generate the parameters of a large neural network from a low-dimensional latent representation. Given an architecture $\Phi$ with $P_\Phi$ parameter slots, we write $\boldsymbol{\theta}_f=\mathcal{G}(\boldsymbol{\xi}_f)...

📖 Read original article


185. Constant Individual Regret in General Games ​

Author: Mingyang Liu, Gabriele Farina, Asuman Ozdaglar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.GT

arXiv:2608.31166v1 Announce Type: new Abstract: Uncoupled no-regret dynamics provide a decentralized route to equilibrium, but prior guarantees for individual regret retain a polylogarithmic dependence on the horizon. We remove this dependence for every finite $N$-player normal-form game under full-...

📖 Read original article


186. Fine-Tuning Qwen3-27B for C-to-Rust Code Translation: A Three-Stage Curriculum of Pretraining, Debugging-Aware SFT, and Task-Specific SFT ​

Author: Pu Zhao, Changdi Yang, Yixiao Chen, Yi Gao, Yifan Cao, Haochen Zeng, Yanzhi Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.ET, cs.LG

arXiv:2608.13681v1 Announce Type: cross Abstract: Translating C code into safe, idiomatic Rust is a longstanding software-engineering goal because it can eliminate entire classes of memory-safety vulnerabilities while preserving the functional behavior of legacy systems. Large language models (LLMs)...

📖 Read original article


187. Privacy-Preserving Detection of Rare Disease-Associated Cell Subsets via Secure Multi-Party Computation ​

Author: \c{S}. Selcan Magara, Esther Havemann, Debora Jutz, Ali Burak "Unal, Mete Akg"un
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.20118v1 Announce Type: cross Abstract: The detection of rare disease-associated cell subsets from high-dimensional single-cell measurements is critical for understanding diseases such as leukaemia and viral infections. CellCnn, a convolutional neural network (CNN) designed for this task, ...

📖 Read original article


188. Propensity Straight-Through Gradients for Discrete Stochastic Systems ​

Author: Jose M. G. Vilar, Leonor Saiz
Published: 9/1/2026, 4:00:00 AM
Categories: q-bio.QM, cond-mat.dis-nn, cs.LG, physics.comp-ph, q-bio.MN

arXiv:2608.25631v1 Announce Type: cross Abstract: Continuous-time Markov chains (CTMCs) provide the backbone for modeling discrete stochastic dynamics across applied, physical, and biological sciences. Their integration with modern gradient-based machine learning, however, is limited by the hard cat...

📖 Read original article


189. The Signal in the Noise: An Auditable Reliability Layer for Biomedical Text Classification ​

Author: Moustafa Yehia Hassan, Sharon Wong, Woh Kai Xuan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.28595v1 Announce Type: cross Abstract: Biomedical NLP pipelines routinely presuppose clean input text, yet large-scale corpora assembled through automated PDF parsing harbour pervasive OCR-like artifacts, token splits and merges, hyphenation remnants, and character-level corruption, that ...

📖 Read original article


190. Preference Elicitation for Policy Optimization and Application to Aligning Heart Transplantation with Human Values ​

Author: Itai Zilberstein, Ioannis Anagnostides, Zachary W Sollie, Arman Kilic, Tuomas Sandholm
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.28620v1 Announce Type: cross Abstract: Preference elicitation is essential for aligning AI systems with human values. Prior approaches (e.g., for organ allocation) often ask stakeholders to compare the decisions of an algorithm (e.g., patient A vs. patient B). Such a decision-level approa...

📖 Read original article


191. Machine Learning-Enhanced Tabu Search for Tactical Wireless Network Design ​

Author: Wissem Ahmed Zaid, Alain Hertz, Defeng Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, math.CO

arXiv:2608.28627v1 Announce Type: cross Abstract: Designing high-performance tactical wireless networks under realistic operational constraints gives rise to challenging combinatorial optimization problems, where the evaluation of candidate solutions relies on detailed physical and traffic-aware mod...

📖 Read original article


192. AutoScientist-Quant: Self-Evolving Coding Agents for Automatic Research in Quantitative Investment ​

Author: Zongqian Li, Yaoyiran Li, Yaohui Guo, Ming Zhang, Nigel Collier, Eugene Ie
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.28632v2 Announce Type: cross Abstract: Large language model agents can discover alphas, yet current methods have three weaknesses. The search cannot adapt during the run, automation usually ends at alpha generation while library selection and model choice stay manual, and alpha discovery ...

📖 Read original article


193. Reward-Oracle MCTS for Formal Theorem Proving: Sample-Efficient Search and the Need for Kernel-Level Proof Auditing ​

Author: Bodla Krishna Vamshi, Haizhao Yang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.LO

arXiv:2608.28639v1 Announce Type: cross Abstract: Formal theorem proving with large language models remains challenging due to the difficulty of navigating large proof search spaces efficiently. Existing tree search approaches either feed verbose compiler error messages directly into the generation ...

📖 Read original article


194. From Extraction to Governed Memory: Multi-Agent Knowledge Graph Construction with Domain-Expert Review ​

Author: Pranav Bykampadi, Neel Mokaria, Vishesh Narayan, Faizan Wajid, Ashok Agrawala
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CE, cs.DL, cs.ET, cs.LG

arXiv:2608.28642v1 Announce Type: cross Abstract: Knowledge graphs used by agentic systems are often treated as flat stores of extracted triples, with little record of who owns a fact, why it was admitted, or how it should be used downstream. We argue that reliable agentic knowledge systems require ...

📖 Read original article


195. How Language Models Choose Sides: Internal Representations of Instruction Hierarchy ​

Author: Enrique Balp-Straffon, Chih-Hao Hsu, Rushiraj Gadhvi, Sunishchal Dev, Callum Stuart McDougall, Anusha Mujumdar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.28648v1 Announce Type: cross Abstract: We study how instruction-tuned LLMs arbitrate direct conflicts between system and user instructions. We introduce a benchmark of 41 paired constraints with deterministic verifiers and evaluate eight models under matched baseline, conflict, and same-c...

📖 Read original article


196. Test-Time Scaling for Scientific Equation Discovery ​

Author: Haowei Lin, Hubert Lim, Xiangyu Wang, Letian Huang, Di He
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.28660v1 Announce Type: cross Abstract: Test-time scaling (TTS) improves language model reasoning by allocating additional test-time compute, but prior work mainly studies closed-ended tasks such as math and coding. We study TTS for automated equation discovery, an open-ended setting where...

📖 Read original article


197. FrameScope: Temporal Data Valuation for Stream Active Learning in Autonomous Vehicle Systems ​

Author: Yuheng Zhu, Man-Ki Yoon
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.28672v1 Announce Type: cross Abstract: Autonomous vehicles operate in dynamic, ever-changing environments where new scenarios and edge cases constantly emerge. As a result, static learning models are inadequate for ensuring safe and reliable operation. Continuous learning is essential for...

📖 Read original article


198. AdaptAV: Continuous Adaption of Vision Models for Autonomous Vehicles Using Cloud-based Oracle ​

Author: Yuheng Zhu, Dhruva Ungrupulithaya, Boluo Ge, Man-Ki Yoon
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.28673v1 Announce Type: cross Abstract: Deploying vision perception models in autonomous vehicles requires that we prioritize inference speeds, resulting in a model with shallower architectures and lesser model parameters (i.e., more pruned). Such small models do not generalize well, which...

📖 Read original article


199. Distributed Semantic Segmentation With Improved Rate-Distortion Trade-Off ​

Author: Danish Nazir, Timo Bartels, Thorsten Bagdonat, Tim Fingscheidt
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.28684v1 Announce Type: cross Abstract: Distributed deep neural networks (DNNs) for dense perception tasks such as semantic segmentation execute an encoder DNN on edge devices, and a decoder DNN typically on a large-scale cloud platform with a particular constraint on transmission bitrate....

📖 Read original article


200. Data Diversity, Not Frequency Invariance: A Controlled and Self-Audited Study of Compression-Robust Deepfake Detection ​

Author: Abbas Aliyev, Samir Rustamov
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.MM

arXiv:2608.28685v1 Announce Type: cross Abstract: Frequency features and compression-invariant representation learning are widely assumed to be key to deepfake detection that survives video compression. We test this with CAFRL - block-DCT and FFT-phase streams, compression-level-conditioned band att...

📖 Read original article


201. ORDDAR: Observation-Driven Reasoning for Distortion-Resilient Decision, Action, and Cognitive Recovery ​

Author: Deblina Kar, Anant Nawalgaria, Shyamal Kumar Das Mandal
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.28704v1 Announce Type: cross Abstract: AI agents increasingly perform long-term reasoning, planning, tool use, memory integration, and autonomous decision making, yet erroneous intermediate states can propagate and cause inconsistent decisions and unreliable outputs. Existing reasoning ap...

📖 Read original article


202. Generation of High-Level Concepts in 3D Scene Graphs via Autoregressive Diffusion ​

Author: Jose Andres Millan-Romera, Samuel Cognolato, Holger Voos, Jose Luis Sanchez-Lopez, Luciano Serafini
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.28733v1 Announce Type: cross Abstract: Indoor 3D Scene Graphs (3DSGs) represent environments as multi-layer hierarchies that connect observed geometric primitives (e.g., planes) to higher-level metric-semantic concepts (e.g., rooms, floors, buildings), enabling incremental spatial reasoni...

📖 Read original article


203. Adversarial Calibration Attack on Autonomous Vehicles ​

Author: Liangkai Liu, Qingzhao Zhang, Kang G. Shin
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.ET, cs.LG, cs.SY, eess.SY

arXiv:2608.28778v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) rely on accurate camera-LiDAR calibration for multimodal sensor fusion. In practice, calibration can drift due to vibration, temperature variation, or minor sensor displacement, motivating online calibration algorithms that ...

📖 Read original article


204. ASTRA - Agentic System for Ticket Resolution and Analysis ​

Author: Shashidhar Reddy Javaji, Mohamed Trabelsi, Jin Cao, Huseyin Uzunalioglu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.IR, cs.LG, cs.SE

arXiv:2608.28790v1 Announce Type: cross Abstract: Technical operations teams resolve large volumes of incidents by synthesizing fragmented evidence from ticket text, historical cases, system logs, and technical documentation. Existing automation often relies on monolithic generation without explicit...

📖 Read original article


205. Separable Nonnegative Matrix Factorization Using Powered Ratio-of-Norms Regularization ​

Author: Matthew McCarver, Jing Qin
Published: 9/1/2026, 4:00:00 AM
Categories: math.NA, cs.LG, cs.NA, math.OC

arXiv:2608.28799v1 Announce Type: cross Abstract: Separable nonnegative matrix factorization (SNMF) has been widely used for low-rank representation and clustering of nonnegative data, owing to its ability to produce part-based and interpretable decompositions. In particular, SNMF is closely related...

📖 Read original article


206. Quantitative Target Convergence and Uniform-in-Time Propagation of Chaos for Langevin-Regularized SVGD ​

Author: Sayan Banerjee, Dohyeon Kim
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.PR

arXiv:2608.28827v1 Announce Type: cross Abstract: We establish quantitative convergence to the target and uniform-in-time propagation of chaos for Langevin-regularized Stein variational gradient descent. The Stein interaction need not be small relative to the confining Langevin drift and does not ge...

📖 Read original article


207. Representation Learning with Quantum Signal Processing ​

Author: Junqi Wang, Junyu Liu
Published: 9/1/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.LG, stat.ML

arXiv:2608.28828v1 Announce Type: cross Abstract: Representation learning begins when training changes the features that define similarity between data. A frozen-kernel model only reweights a fixed geometry. We establish quantum signal processing (QSP) as a solvable quantum model of the representati...

📖 Read original article


208. Toward Postural State Classification in Immersive VR with Multimodal Data and Explainability Analysis ​

Author: Nipa Anjum, Md Irfan Pavel, Robert Gonzalez Jr, Kevin Desai, Alberto Cordova, M. Rasel Mahmud, John Quarles
Published: 9/1/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.LG

arXiv:2608.28844v1 Announce Type: cross Abstract: Ensuring a safe virtual reality (VR) experience requires systems that can predict and respond when users lose their balance. Although prior work has examined fall prediction and motion sickness, many approaches are regression-based and postural state...

📖 Read original article


209. A rigor-matched audit of periodic-step layer skipping for efficient llm inference: conflayers versus swift, with a supplemental analysis of trained routing alternatives ​

Author: Prateek Kumar Sikdar, Arpan Ghosh
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.28846v1 Announce Type: cross Abstract: Layer-skipping methods for efficient LLM inference decide, at some granularity, which transformer layers to execute for a given input. We present a rigor-matched, three-seed audit of two periodic-step, search-based methods that make this decision onl...

📖 Read original article


210. Uncertainty-Aware Multi-Task Learning for Joint Modulation Recognition and SINR Estimation ​

Author: Kosar Nourolahi, Vahid Ghasemi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2608.28865v1 Announce Type: cross Abstract: Joint modulation recognition and signal-to-interference-plus-noise ratio (SINR) estimation can reduce duplicated processing in intelligent receivers, but the two tasks have different uncertainty characteristics. This letter proposes an uncertainty-aw...

📖 Read original article


211. Generative Translation Priors: Bayesian Imaging with Cross-Modality Image Translation ​

Author: Evan Bell, Jiaming Liu, Yifan Chen, Yu Sun
Published: 9/1/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG

arXiv:2608.28872v1 Announce Type: cross Abstract: The ability to leverage images from co-available modalities to inform target-domain reconstruction is highly desirable in imaging algorithms. In this work, we introduce Generative Translation Priors (GTP)--a Bayesian framework that transforms diffusi...

📖 Read original article


212. MWIR-4-Plastic: The Identification of Complex End-of-Life Industrial Plastic using Mid-wave Infrared Hyperspectral Imaging and Machine Learning ​

Author: Elias Arbash, Andr'ea de Lima Ribeiro, Filipa Sim~oes, Ahmed Jamal Afifi, Aldino Rizaldy, Yuleika Madriz, Samuel Thiele, Sandra Lorenz, Margret Fuchs, Pedram Ghamisi, Paul Scheunders, Richard Gloaguen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.28874v1 Announce Type: cross Abstract: The automated sorting of shredded black plastics from end-of-life (EOF) industrial waste presents a significant challenge in recycling facilities, primarily due to the limitations of current sensing and analytical approaches. Existing studies predomi...

📖 Read original article


213. Moving the Mean Toward the Known Good, Not Beyond It: What Inference-Time Interventions and Weight Consolidation Buy in Open-Ended Generation ​

Author: Roberto I. Ono Filho
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.28886v1 Announce Type: cross Abstract: What does a generation loop gain from learning on its own verified successes? In cycles of generate, verify, select and LoRA-consolidate on online bin packing, training on value-filtered candidates shifts what the model writes on held-out variants to...

📖 Read original article


214. mmIR: Frequency-Space Inverse Rendering for 3D Millimeter-Wave Radar ADC Synthesis ​

Author: Adnan Armouti, Yixuan Gao, Rajalakshmi Nandakumar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.GR, cs.LG, eess.SP

arXiv:2608.28913v1 Announce Type: cross Abstract: High-resolution 3D radar data is scarce. Commodity mmWave sensors use small antenna arrays that limit angular resolution to several degrees, and existing datasets provide only 2D range-azimuth maps or sparse point clouds rather than raw analog-to-dig...

📖 Read original article


215. Leveraging Turn-taking Dynamics for Intent Recognition in Multi-party Conversations ​

Author: Galo Castillo-L'opez, Alexis Lombard, Ga"el de Chalendar, Nasredine Semmar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.28926v1 Announce Type: cross Abstract: We propose a multi-task learning approach for multi-party dialogue intent recognition that leverages an auxiliary task that models turn-taking dynamics. Specifically, we introduce turn-transition entropy, a self-supervised target computed from the se...

📖 Read original article


216. The Hallucination Signal Is a Mean Shift: Why Simple Probes Suffice ​

Author: Jungseob Lee, Jaehyung Seo, Heuiseok Lim
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.28930v1 Announce Type: cross Abstract: Hidden-state probes effectively detect LLM hallucinations, but the geometry of the signal remains poorly characterized, driving increasingly complex probe architectures. Across three 7B-scale models and three datasets in a paired-example paradigm, we...

📖 Read original article


217. MERIT: Mitigating Exposure Bias in Generative XMC for User-Interest Propensity Modeling ​

Author: Abhinav Mahajan, Arindam Sarkar, Prakash Mandayam Comar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.IR, cs.LG

arXiv:2608.28931v1 Announce Type: cross Abstract: Matching users to interest categories at scale is central to personalized shopping, but the task is challenging in large e-commerce platforms, where label spaces continually evolve and user-interest signals are sparse and long-tailed. Autoregressive ...

📖 Read original article


218. Oculi: A Conversational Agentic Platform for Automated Credit Risk Analysis ​

Author: Vennise Ho, Kristian Diana, Sandy Mourad, Milena Pilipovic, Vineel Nagisetty, Hossein Hajimirsadeghi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.28944v1 Announce Type: cross Abstract: Credit risk analysis in financial institutions traditionally requires analysts to manually write SQL queries, run statistical computations, and build visualization dashboards. This is a time-consuming workflow that limits exploration to familiar segm...

📖 Read original article


219. The information geometry of product-reference discrete diffusion: Interaction growth complexity and optimal scheduling ​

Author: Martin J. Wainwright
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, math.ST, stat.TH

arXiv:2608.28949v1 Announce Type: cross Abstract: We study a class of product-reference diffusion algorithms for sampling from a discrete distribution. We show that their sampling performance can be characterized using a path-based measure of data geometry that we call the interaction growth complex...

📖 Read original article


Author: Yanbo Li, Chujie Zheng, Jiahao Xu, Chetan Bhole, Lingyu Zhang, Puneet Singh Ahluwalia, Kevin Nguyen, Raghavan Muthuregunathan, Santhosh Sachindran, Sachin Ahuja, Fedor Borisyuk
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.28965v1 Announce Type: cross Abstract: People search must map free-form location phrases to geographic entities used as structured retrieval filters. Lexical standardizers handle canonical names well but are brittle to aliases, misspellings, metropolitan expressions, and same-name ambigui...

📖 Read original article


221. Brain-Language-Action (BLA) Models: Language-Conditioned EEG for Robotics Control ​

Author: Alexandr Plashchinsky
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.28967v1 Announce Type: cross Abstract: Electroencephalography (EEG)-based robotic control is commonly formulated as a direct classification problem, in which electrical neural signals are mapped to a fixed set of discrete actions. However, the limited separability and high noise of EEG si...

📖 Read original article


Author: Dhritiman Das, Chujie Zheng, Ronak Kaoshik, Pratik Dixit, Vishal Shah, Yanbo Li, Jiahao Xu, Manika Agarwal, Chinmay Naik, Lingyu Zhang, Chetan Bhole, Chirag Bhanuprasad Mehta, Meng Zheng, Puneet Singh Ahluwalia, Shirisha Singh, Ping Jin, Manas Apte, Gokulraj Mohanasundaram, Tugrul Bingol, Raghavan Muthuregunathan, Fedor Borisyuk
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.28968v1 Announce Type: cross Abstract: Semantic Search on LinkedIn must retrieve relevant profiles from a corpus of hundreds of millions in response to natural-language queries such as "a fintech founder in Berlin who worked in payments." The deployed relevance policy is bottleneck-orient...

📖 Read original article


223. The Illusion of Replacement: Rethinking Specialized Machine Learning Models in the Foundation Model Era ​

Author: Kiyan Rezaee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.28980v1 Announce Type: cross Abstract: Can the specialized architectures that machine learning has traditionally built for structured data be replaced by language-based models? This question is examined through a review of 159 papers (2016--2026) across nine modalities, with predictive ac...

📖 Read original article


224. Jigsaw-CRL: Recovering Global Latent Causal Order from Fragmented Multi-Client Interventions ​

Author: Haijie Xu, Chen Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.28991v1 Announce Type: cross Abstract: Causal representation learning (CRL) aims to recover latent causal variables and their structural relations from high-dimensional observations. Existing CRL methods typically assume that all environments are defined over the same latent variables, or...

📖 Read original article


225. Sharp Restricted Isometry Thresholds for Global Minima of Rank-Restricted Matrix LASSO ​

Author: Richard Y. Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.IT, cs.LG, math.IT, math.OC, math.ST, stat.TH

arXiv:2608.29018v1 Announce Type: cross Abstract: We determine the sharp restricted isometry threshold for recovery at global minima of the rank-restricted matrix LASSO. For target rank $r_{\star}$, if the rank-$k$ RIP constant satisfies $\delta<\delta_{\mathrm{sharp}}(k/r_{\star})$, where $\delta_{...

📖 Read original article


226. Spectral-Embedded Operator Learning for Three-Phase Interfacial Flow: A Ternary Cahn-Hilliard-Navier-Stokes Benchmark ​

Author: Muhammad Abid, Arth Sojitra, Omer San
Published: 9/1/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG

arXiv:2608.29069v1 Announce Type: cross Abstract: Operator-learning surrogates have been benchmarked largely on single-field, single-interface problems, leaving unclear whether architectural choices validated in those settings transfer to constrained, multiphase flows. We introduce a three-phase int...

📖 Read original article


227. Optimally Selecting Representative Agents from a Metric Space ​

Author: Benjamin Cookson, Eva Deltl, Yeeseok Oh
Published: 9/1/2026, 4:00:00 AM
Categories: cs.GT, cs.LG

arXiv:2608.29097v1 Announce Type: cross Abstract: This paper studies the problem of proportionally fair clustering, where the goal is to select $k$ ``centers'' from a metric space that fairly represent a set of agents who also lie in the metric space. Specifically, we focus on finding a clustering s...

📖 Read original article


228. Clustering as Approximation by Constrained Projectors: Theory and Guarantees ​

Author: Angshul Majumdar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.29102v1 Announce Type: cross Abstract: This paper develops a unified theoretical framework showing that a broad family of clustering methods, including k-means, fuzzy c-means, kernel k-means, kernel FCM, and spectral clustering, can all be expressed as structured low-rank projectors actin...

📖 Read original article


229. Emergent Misalignment Is Not Magical ​

Author: Mingxuan Li, Qirun Dai, Heran Wang, Chenhao Tan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.29118v1 Announce Type: cross Abstract: Fine-tuning large language models (LLMs) on narrowly harmful datasets can lead to misalignment broadly, a phenomenon known as emergent misalignment (EM). EM poses a challenge for AI safety and our understanding of LLMs. Prior work often frames EM as ...

📖 Read original article


230. APIFlow-Bench: Measuring Whether Agents Survive Long, Dependent API Workflows ​

Author: Zelin Wan, Arash Nourian, Xiaoxiao Li, Nihar Nandan, Kamalakannan Nandagopal
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE

arXiv:2608.29128v1 Announce Type: cross Abstract: Tool-using agents are commonly evaluated by a single bit: whether an end-to-end workflow completed. This metric fails to distinguish failures that matter in production, such as expired credentials, malformed payloads, or correct execution followed by...

📖 Read original article


231. Uniform Statistical Convergence of Empirical Sinkhorn Potentials with Exponential and Polynomial Dependence on the Regularization Parameter ​

Author: Denis Belomestny
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.OC

arXiv:2608.29152v1 Announce Type: cross Abstract: We study the empirical Sinkhorn estimator of the entropic optimal transport potentials under the uniform loss. Since the potentials are only unique up to additive constants, we measure the error using the quotient supremum norm, defined as $d_\infty(...

📖 Read original article


232. Subtraction-Based Tumor Segmentation and Lesion-Centered pCR Prediction for the MAMA-MIA Challenge ​

Author: Kai Geissler, Raphael Sch"afer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.29162v1 Announce Type: cross Abstract: We describe the submission of team FME to the MAMA-MIA Challenge, which evaluated primary tumor segmentation and prediction of pathological complete response (pCR) from pretreatment dynamic contrast-enhanced breast MRI on an external multi-country co...

📖 Read original article


233. Hyper-Fold: Exploring the Expressive Limit of Sequence-Geometry Learning for Proteins via Hypergraph Modeling ​

Author: Yifan Feng, Guanjie Cheng, Shihui Ying, Shaoyi Du, Yue Gao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.29207v2 Announce Type: cross Abstract: Protein structure modeling rests on a single computational primitive: the interaction between what a residue is (sequence content) and where it sits (three-dimensional geometry). What is the expressive limit of this layer class? We show that the comp...

📖 Read original article


234. AdaVLA: Adaptive Step Flow Matching for Training-free Acceleration of Vision-Language-Action Models ​

Author: Sunghwan Han, Youngtae Han, Youngmin Yi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2608.29208v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models, built upon Vision-Language Models (VLMs), have significantly enhanced robotic capabilities by leveraging internet-scale knowledge and multimodal reasoning. However, the intensive computational overhead of VLAs con...

📖 Read original article


235. Validating FKG.in: Soundness Assessment in LLM-Augmented Indian Food Knowledge ​

Author: Saransh Kumar Gupta, Armaan Shah, Lipika Dey, Partha Pratim Das, Ramesh Jain
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.IR, cs.LG

arXiv:2608.29249v2 Announce Type: cross Abstract: The online culinary ecosystem is increasingly populated by recipe content generated, modified, or summarized by Large Language Models (LLMs). While often plausible, such outputs may contain hallucinated ingredients, misrepresented quantities, or cult...

📖 Read original article


236. QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation ​

Author: Yaroslav Prytula, Anton Popov, Dmytro Fishman
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.29253v1 Announce Type: cross Abstract: Instance segmentation of overlapping cells in microscopy remains challenging due to semi-transparent structures that produce weak boundaries and mixed visual evidence in overlap regions. Existing methods address this through local regions of interest...

📖 Read original article


237. A-MADiff: Attention-Guided Multi-Agent DRL with Diffusion Policies for Memory-Aware Task Orchestration in Mobile AIGC Networks ​

Author: Chongzhi Wu, Zhengtao Li, Jiawen Kang, Jinbo Wen, Xiaohuan Li, Maomao Zhang, Ekram Hossain
Published: 9/1/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2608.29255v1 Announce Type: cross Abstract: Artificial Intelligence-Generated Content (AIGC) services employ Generative AI (GenAI) models to automatically generate diverse content. Mobile AIGC networks host GenAI models on edge-located AIGC Service Providers (ASPs) to deliver low-latency and p...

📖 Read original article


238. Sense Once, Serve Many: Common-Trace Factorized Constrained PPO for Online Sensing-Session Consolidation in Multi-Tenant ISAC Networks ​

Author: Dang-Dung Vu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.NI, cs.LG

arXiv:2608.29256v1 Announce Type: cross Abstract: Integrated sensing and communication (ISAC) networks can serve compatible requests through shared sensing sessions, but consolidation couples admission, reuse, profile selection, sensing service-level agreements (SLAs), communication quality of servi...

📖 Read original article


239. Signed random Fourier features for fast density estimation with indefinite kernels ​

Author: Xie Wang, Nicolas Langren'e, Wen Chen
Published: 9/1/2026, 4:00:00 AM
Categories: stat.CO, cs.LG, math.PR, stat.ML

arXiv:2608.29265v1 Announce Type: cross Abstract: Kernel density estimation (KDE) is one of the most fundamental statistical estimators of density functions. Its direct implementation on a dataset of $N$ points incurs an $\mathcal{O}(N^{2})$ computational cost, which is prohibitive for large-scale d...

📖 Read original article


240. Learning Simple Test-Time Environments for LLM Web Agents ​

Author: Junxuan Li, Zijun Liu, Ziyi Huang, Peng Li, Yuzhou Liu, Ming Yan, Yang Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.29305v1 Announce Type: cross Abstract: Large language model (LLM) agents have demonstrated remarkable proficiency in manually constructed environments, yet their performance frequently collapses when transitioned to complex real-world settings. Existing research largely attribute this deg...

📖 Read original article


241. APPSolver: Adaptive Patch Partitioning for Point-Wise Ship Flow Prediction on Unstructured Meshes ​

Author: Wenhua Huo, Fenglei Han, Wangyuan Zhao, Xiao Peng, Chunhui Wang, Jialin Wu, Jiayi Han
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.29355v1 Announce Type: cross Abstract: Large non-uniform point sets make direct attention-based surrogate modeling costly for ship hydrodynamics. We introduce APPSolver, a point-wise flow-prediction framework built around Adaptive Patch Partitioning (APP), a deterministic quadtree represe...

📖 Read original article


242. Spectral Analysis for Sparse Matrix Computation: Insights and Potential ​

Author: Ruifeng Zhang, Xipeng Shen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.PF, cs.LG

arXiv:2608.29362v1 Announce Type: cross Abstract: Sparse computations are fundamental to scientific computing, graph analytics, and machine learning, yet their performance is highly sensitive to the diverse sparsity and patterns. This is because cache reuse, memory coalescing, and load balancing dep...

📖 Read original article


243. Evaluating Tiny Recursive Models Across Training for Code Generation ​

Author: Anjani Sirivella, Aanisha Newaz, Glaucia Melo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE

arXiv:2608.29376v1 Announce Type: cross Abstract: Code generation increasingly relies on large transformer models, whose capability advances with scale. Yet such a scale is costly, creating demand for small models, especially where data is limited. Recursive models address this by reusing a single b...

📖 Read original article


244. FiLM-GPNet: Geometry-Aware Pseudo-Supervised Phase Restoration with Zero-Shot Generalization for Large Temporal InSAR Stacks ​

Author: Getnet Demil, Muhammad Farhan Humayun, Tomi Westerlund, Jukka Heikkonen, Mourad Oussalah
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.29384v1 Announce Type: cross Abstract: The growing availability of dense commercial Synthetic Aperture Radar (SAR) time series enables temporal Interferometric SAR (InSAR) analysis, but fixed classical filters fail under heterogeneous acquisition geometries, degrading phase quality and te...

📖 Read original article


245. Explanations, Prompts, and Formalizations: Arguments for New Norms in LLM-Enabled Mathematical Research ​

Author: Axel Boldt
Published: 9/1/2026, 4:00:00 AM
Categories: math.HO, cs.LG

arXiv:2608.29401v1 Announce Type: cross Abstract: As several mathematical conjectures have recently been settled using large language models (LLMs), the mathematical community has formulated norms and recommendations regarding the publishing of such results. These norms do not cover the disclosure o...

📖 Read original article


246. A Visual Question Answering Model to Automate Nondestructive Evaluation Image Analysis ​

Author: Mehrdad Shafiei Dizaji, Hoda Azari
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.29408v1 Announce Type: cross Abstract: This study introduces a Visual Question Answering model designed specifically for nondestructive evaluation applications. VQA models allow inspectors to interactively query NDE images, asking targeted questions like, Is there a crack or Where is the ...

📖 Read original article


247. Polis: 3D Self-Supervision at City Scale ​

Author: Alexander Rusnak, Sophia Kovalenko, Jingru Wang, Ismail Moudden, Xiru Wang, Fr'ed'eric Kaplan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.29426v1 Announce Type: cross Abstract: Reliable semantic representations derived from city-scale 3D models are increasingly important for urban analysis, infrastructure monitoring, autonomous systems, and heritage conservation. However, urban scenes of large spatial extent captured throug...

📖 Read original article


248. Content Exploration Beyond the Feed: Creator Supply and the Shared Corpus ​

Author: Yuanyuan Shen, Yiren Yan, Wenjie Li, Chunhui Zhu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.IR, cs.LG, econ.EM, stat.AP, stat.ML

arXiv:2608.29430v1 Announce Type: cross Abstract: Industrial recommenders give new content initial views through budgeted exploration, then use early performance to decide further delivery. On many short-video platforms, exploration is the primary way new videos reach viewers. Viewer-side tests meas...

📖 Read original article


249. Item-Mean Surrogates: Why Richer Persona Data Fail to Improve LLMs as Human Surrogates ​

Author: Daehwan Ahn, Chengfeng Mao, Dokyun Lee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.CY, cs.LG

arXiv:2608.29455v1 Announce Type: cross Abstract: LLMs are increasingly used as human surrogates, often on the premise that richer persona data could make them substitutes or exploratory tools for specific individuals. We test this premise across four datasets covering more than 400,000 participants...

📖 Read original article


250. Benchmark Contamination: A Taxonomy Organized by Defeated Mitigation ​

Author: Johanna Angulo, V'ictor Yeste, Hector Espinos-Morato
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG

arXiv:2608.29463v1 Announce Type: cross Abstract: A benchmark score is a joint property of the model, the evaluation harness, the elicitation budget, the sampled population, and contamination status. Leaderboards publish the model and the score, so capability and leakage stay observationally equival...

📖 Read original article


251. Deciding When to Decide: Testing Operational Suboptimality Under Distributional Shift ​

Author: Minxing Zheng, Holly Wiberg, Shixiang Zhu
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2608.29465v1 Announce Type: cross Abstract: Deployed decisions are often optimized once and retained because updates impose operational, regulatory, or switching costs. As operating conditions change, when should such decisions be re-optimized? We study this question for stochastic optimizatio...

📖 Read original article


252. ARMOR: Manifold-Oriented Training for Adversarially Robust Aerial Object Detection under Data Scarcity ​

Author: Haoran Wang, Matthew Lau, Alec Helbling, Matthew Hull, ShengYun Peng, Mansi Phute, Martin Andreoni, Willian T. Lunardi, Duen Horng Chau, Wenke Lee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.CR, cs.LG

arXiv:2608.29510v1 Announce Type: cross Abstract: Aerial object detection is increasingly deployed in real-world applications, but models remain vulnerable to physical, universal adversarial patches that cause them to miss objects. Furthermore, defenders face the practical constraint of training dat...

📖 Read original article


253. LoGo: Token-Level Dynamic Local-Global Attention ​

Author: Yuqi Pan, Zheng Li, Bohao Tang, Zhen Qin, Guoqi Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.29539v1 Announce Type: cross Abstract: As context lengths scale, attention increasingly becomes a primary computational bottleneck in large language models. Standard Transformers remain powerful but computationally inefficient, as they allocate the same attention budget to every token reg...

📖 Read original article


254. TACS: Trajectory-Aware Candidate Selection for LLM Jailbreak Suffix Optimization ​

Author: Shiliang Xiao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.29564v2 Announce Type: cross Abstract: Gradient-based jailbreak suffix optimization methods typically update the suffix by retaining the candidate with the lowest current loss. We show that this seemingly natural design is fundamentally myopic: candidates that look better under the curren...

📖 Read original article


255. Towards a Systems Foundation for Agentic Skills: Architecture, Lifecycle, and Security ​

Author: Sanket Badhe, Deep Shah, Priyanka Tiwari, Nehal Kathrotia
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.MA

arXiv:2608.29596v1 Announce Type: cross Abstract: Autonomous large language model (LLM) agents increasingly face reliability, context consumption, and execution stability bottlenecks when deployed on complex, long-horizon tasks. While monolithic prompt engineering and stateless tool-calling paradigm...

📖 Read original article


256. $\mathcal{N}_0$-Foundation: Towards the Age of Tactile Intelligence ​

Author: NeoteAI Team, Fudan TEAI Team
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.CV, cs.LG

arXiv:2608.29601v1 Announce Type: cross Abstract: We present $\mathcal{N}_0$-Foundation, a paradigm for tactile-enabled embodied manipulation, which integrates tactile sensing hardware, large-scale multimodal data, tactile representation learning, and standardized evaluation. First, we engineer the ...

📖 Read original article


257. Cross-lingual Functional Vectors for Emotion Detection in Large Language Models ​

Author: Jieying Xue, Phuong Minh Nguyen, Minh Le Nguyen, Shogo Okada
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.29613v1 Announce Type: cross Abstract: Function vectors (FVs) have recently emerged as a promising mechanism for steering the behavior of large language models (LLMs) by injecting task-specific latent direction representations derived from in-context demonstrations. While prior studies ha...

📖 Read original article


258. Forward-Deployed Full-Stack Engineering for Autonomous Cloud MLOps ​

Author: Sagar Srinivas Sakhinana, Venkataramana Runkana
Published: 9/1/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.LG

arXiv:2608.29615v1 Announce Type: cross Abstract: Across industries, machine-learning systems support applications ranging from prediction and anomaly detection to forecasting, optimization, and scheduling, yet operationalizing these systems requires coordinating application development, model pipel...

📖 Read original article


259. ButterMamba: Butterworth-Enhanced Spatial-Temporal Mamba for Efficient Traffic Flow Prediction ​

Author: Limiao Zhang, Yuhui Lu, Jie Gao, Hao Jiang, Haiping Ma, Xingyi Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: physics.soc-ph, cs.LG, physics.data-an

arXiv:2608.29658v1 Announce Type: cross Abstract: Accurate traffic flow prediction is fundamental to intelligent transportation systems, playing a pivotal role in urban mobility optimization and smart city development. While Graph Neural Networks (GNNs) integrated with time series forecasting have e...

📖 Read original article


260. Transformer-Based Flow Shop Scheduling Using MILP-Generated Training Data ​

Author: Roderich Wallrath
Published: 9/1/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2608.29690v1 Announce Type: cross Abstract: Advances in machine learning (ML) have created new opportunities to complement traditional operations research (OR) methods. In particular, transformer models can capture complex interactions in token sequences by mapping tokens into a high-dimension...

📖 Read original article


261. Neural ODE enhanced linear mixed effect models for estimating complex association patterns of time-varying covariates with the marker trajectory ​

Author: Zhe Aurore Li, Quentin Clairon, C'ecilia Samieri, Rodolphe Thi'ebaut, M'elanie Prague, C'ecile Proust-Lima
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.29714v1 Announce Type: cross Abstract: Longitudinal cohort studies produce repeated data that enable the assessment of time-varying association patterns between exposures and health outcomes. Classical linear mixed-effects models (LMMs) can accommodate a large variety of association patte...

📖 Read original article


262. A Unified Perspective on Conformal Prediction and Wasserstein Distributionally Robust Optimization for Uncertainty Quantification ​

Author: Kehan Long, Yiqi Zhao, Pol Mestres, Lars Lindemann, Nikolay Atanasov, Jorge Cort'es
Published: 9/1/2026, 4:00:00 AM
Categories: math.OC, cs.LG, cs.SY, eess.SY

arXiv:2608.29789v1 Announce Type: cross Abstract: Uncertainty quantification from finite data is central to machine learning, optimization, and automation systems, where decisions must remain reliable under limited samples and test-time distribution shift. Conformal prediction (CP) and distributiona...

📖 Read original article


263. The Price of Intelligence: A Quality-Adjusted Price Index for AI Services ​

Author: Louis Yiven Zhu
Published: 9/1/2026, 4:00:00 AM
Categories: econ.GN, cs.LG, q-fin.EC

arXiv:2608.29843v1 Announce Type: cross Abstract: Posted prices for AI inference have fallen steadily since 2024, yet the measured speed of that fall depends almost entirely on the method of measurement. This paper constructs quality-adjusted price indices for the AI inference market from public dat...

📖 Read original article


264. Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation ​

Author: Run Yang, Runpeng Dai, Jie Sun, Jielei Zhang, Fan Zhou, Hongtu Zhu, Peiyi Li, Longwen Gao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.29846v1 Announce Type: cross Abstract: Sampled-token on-policy distillation (OPD) efficiently transfers capabilities from teacher to student using student-generated tokens, requiring teacher probabilities only for sampled tokens. Yet it frequently suffers from diversity distillation failu...

📖 Read original article


265. On the Instance Hardness as a Decision Criterion in TinyML Systems ​

Author: Tobiasz Puslecki, Krzysztof Walkowiak
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.29913v1 Announce Type: cross Abstract: TinyML includes the implementation of machine learning on devices with limited memory and computing resources. With the development of technology, AI systems continue to scale in terms of size and computational requirements. This forces researchers t...

📖 Read original article


266. Continual Test-Time Adaptation via Entropy Sensitivity-Guidance in Strict Online Setting ​

Author: Chandler Timm C. Doloriel, Yunbei Zhang, Muhammad Salman Siddiqui, Tor Kristian Stevik, Fadi Al Machot, Kristian Hovde Liland, Habib Ullah
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.29920v1 Announce Type: cross Abstract: Test-time adaptation (TTA) promises robustness under distribution shift by updating a pretrained model on unlabeled test data, but strict online TTA with batch size one and no access to source data is especially prone to drift or collapse. We introdu...

📖 Read original article


267. Towards Continual Test-Time Adaptation of Vision-Language Models in Open-Vocabulary Semantic Segmentation ​

Author: Chandler Timm C. Doloriel, Yunbei Zhang, Sarthak Kumar Maharana, Muhammad Salman Siddiqui, Tor Kristian Stevik, Fadi Al Machot, Kristian Hovde Liland, Habib Ullah
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.29923v1 Announce Type: cross Abstract: Open-vocabulary semantic segmentation (OVSS) relies on vision-language alignment to recognize arbitrary text-defined categories, yet this alignment is fragile under continual test-time distribution shift. Our diagnostic analysis reveals that entropy ...

📖 Read original article


268. Hallucination Mitigation for Large Vision-Language Models via Implicit Feature Stabilization ​

Author: Aditi Sarker, Rafi Ibn Sultan, Hui Zhu, Dongxiao Zhu, Prashant Khanduri
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.29924v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) are prone to hallucinations: they fluently describe objects, attributes, and scenes that are not in the image. We connect part of this failure to a measurable property of their representations, feature instability...

📖 Read original article


269. Compression-Aware Abstention: Teaching LLMs to Refuse When KV-Compression Masks Remove Answer Evidence ​

Author: Mohammadali Khodabandehlou, Bhaskar Krishnamachari
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.29934v1 Announce Type: cross Abstract: KV-cache compression reduces LLM inference memory by evicting context tokens, but when the evicted tokens contain answer-bearing evidence, the model may hallucinate instead of recognizing that the compressed context is insufficient. We address this f...

📖 Read original article


270. When Safety Speaks a Language: A Mechanistic Analysis of Safety-Language Identity Entanglement in LLMs ​

Author: Apoorva Upadhyaya, Sandipan Sikdar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.29936v1 Announce Type: cross Abstract: Safety alignment of large language models (LLMs) degrades across languages, yet the internal mechanism driving this asymmetry remains poorly understood. Our work, therefore, presents a systematic mechanistic analysis of multilingual safety using spar...

📖 Read original article


271. Data-Driven Design Optimization of Streaming-Potential-Mediated Electrokinetic Transport of Viscoelastic Fluids in Microchannels ​

Author: Ankan Basu, Sumanta Banerjee
Published: 9/1/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG

arXiv:2608.29939v1 Announce Type: cross Abstract: Streaming-potential-mediated transport of viscoelastic fluids has attracted research attention owing to its applications in electrokinetic energy conversion and microfluidic transport. Existing analytical and semi-analytical models in published liter...

📖 Read original article


272. EDGE: Engine for Deterministic Graph Evaluation through Conversation Simulation from Graph Structured DSL Configuration ​

Author: Ram Kulathumani, Regunathan Radhakrishnan, Anupam Tripathi, Xiangbo Mao, Roshanak Omrani, Keshav Somani, Shwet Kamal Mishra, Shayna Lurya
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.29971v1 Announce Type: cross Abstract: As agentic systems evolve into complex multi agent orchestration workflows, there is a growing and critical need for systematic frameworks that measures an agent's behavioral consistency and determinism. In this paper, we introduce a formal evaluatio...

📖 Read original article


273. An Open-Source, Event-Driven Pipeline for Cryptocurrency Market Data: Ingestion, Forecasting, and On-Chain Fraud Detection ​

Author: Basil Sajid Shaikh, Melrick Mascarenhas, Nuzhat Faiz Shaikh
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.29973v1 Announce Type: cross Abstract: Cryptocurrency markets generate high-frequency, multi-source data that is expensive to work with unless a team already has commercial-grade streaming and warehousing infrastructure in place. This paper describes a fully open-source pipeline that repr...

📖 Read original article


274. Evolutionary Soups: Evolving Mixture-of-Experts for Multi-Objective LLM Alignment ​

Author: Lingxiao Kong, Steffen Staab, Cong Yang, Oya Beyan, Zeyd Boukhers
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG, cs.NE

arXiv:2608.29978v1 Announce Type: cross Abstract: Large language models are increasingly required to generate responses that satisfy multiple competing objectives. Since optimal trade-offs depend on both user preferences and input prompts, controllable multi-objective generation must dynamically ada...

📖 Read original article


275. Partition-Aware Unlearning for Removing Spurious Correlations in Large Vision-Language Models ​

Author: Aditi Sarker, Nazreen Shah, Rafi Ibn Sultan, Rhongho Jang, Dongxiao Zhu, Prashant Khanduri
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.29996v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) achieve strong performance across many multimodal tasks; however, they often exploit spurious object-background correlations, resulting in predictions driven by contextual shortcuts rather than object-relevant vis...

📖 Read original article


276. TEMPO: Temporally-grounded Multi-task Post-training for Large Audio-Language Models ​

Author: Apoorva Kulkarni, Kaousheik Jayakumar, Sreyan Ghosh, Utathya Aich, Ramani Duraiswami, Dinesh Manocha
Published: 9/1/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG

arXiv:2608.29999v1 Announce Type: cross Abstract: Large audio-language models (LALMs) describe audio at the clip level but cannot assign timestamps to the events, speakers, or sounds they identify. Despite being essential for downstream tasks like speech recognition and dense audio captioning, times...

📖 Read original article


277. Beyond Uncertainty: Multi-Solver Disagreement Rewards for Self-Evolving Reasoning Curricula ​

Author: Vinoth Selvendran, Zhanming Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.30035v1 Announce Type: cross Abstract: Self-evolving reasoning frameworks train a Challenger to generate questions exposing a Solver's weaknesses, creating adaptive curricula without human data. However, existing approaches use a single solver's sampling uncertainty as the Challenger's re...

📖 Read original article


278. A Deep Latent Variable Framework for Jointly Modeling Missingness, Measurement Error, and Heterogeneity ​

Author: Yasin Khadem Charvadeh, Grace Y. Yi, Mithat G"onen, Pouya Faroughi
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.30040v1 Announce Type: cross Abstract: Missing data, measurement error, and population heterogeneity are pervasive challenges in analyzing data arising from modern observational studies and machine learning applications. Although these problems frequently coexist and interact, they are of...

📖 Read original article


279. Balance of Benchmarks: Semantic Density Reweighting for Benchmark Multiplicity and Task-Conditioned Evaluation ​

Author: Jhen-Ke Lin
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2608.30044v1 Announce Type: cross Abstract: Language models are commonly compared by averaging scores across a benchmark list with equal weight. Such lists grow through publication outside an explicit measurement design, so equal weighting turns the density of published benchmarks into an impl...

📖 Read original article


280. Mitigating Over-Optimization in PRM-Guided Search in Mathematical Reasoning by Optimizing the Guide ​

Author: Taejong Joo, Diego Klabjan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.30051v1 Announce Type: cross Abstract: Process reward models (PRMs) provide dense step-level guidance for search-based reasoning, enabling inference-time compute to be allocated toward promising partial solutions. However, recent evidence suggests that PRM-guided search can over-optimize ...

📖 Read original article


281. Learning Representations through Token Prediction: Geometry, Approximation, and Downstream Guarantees ​

Author: Shulei Wang
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.ST, stat.TH

arXiv:2608.30072v1 Announce Type: cross Abstract: Token prediction is a central pre-training objective for modern language models. Despite its empirical success, why token prediction learns broadly useful representations remains incompletely understood. We develop a statistical framework connecting ...

📖 Read original article


282. When Does a Classifier Help an LLM? Classifier-Guided Prompting and Hybrid Classifier-LLM Models for Credit-Default Prediction ​

Author: Rishi Datta, Lavanya Prahallad
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.30086v1 Announce Type: cross Abstract: Credit-default prediction is an important task in financial decision making. Traditional methods use fitted classifiers such as logistic regression and random forests on tabular features. Large language models (LLMs) have recently been applied to thi...

📖 Read original article


283. A Hybrid State-Space Approach for Census-Tract Population Estimation ​

Author: Jackson R. Ye, Alexandre V. Morozov
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.30094v1 Announce Type: cross Abstract: Sequence models---the architecture family behind large language models and, increasingly, state-of-the-art image recognition---have redefined how machines learn from high-dimensional data. Yet population estimation from satellite imagery, a task that...

📖 Read original article


284. A Simple Transformer Pipeline for Full-Key Side-Channel Attacks on Uncropped Datasets ​

Author: Jimmy Gammell, Kaushik Roy
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.30105v1 Announce Type: cross Abstract: Deep learning-based side-channel analysis has historically focused on single-byte targets and manually cropped traces, which risks discarding exploitable leakage. While recent work has proposed specialized architectures and resampling techniques to a...

📖 Read original article


285. Aligning Multi-Trajectory Supervision with Policy Optimization for VLA Driving ​

Author: Tian Zhang, Zhuo Huang, Hongrui Ye, Yu Wu, Zengmao Wang, Kaixuan Zhou
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.30122v1 Announce Type: cross Abstract: Vision-language-action (VLA) driving methods increasingly combine multi-trajectory imitation learning with group-relative policy optimization (GRPO), making trajectory selection critical to final performance. However, some high-scoring trajectories t...

📖 Read original article


286. VIBE: Video Instruction-aligned Background music gEneration ​

Author: Aryan Vijay Bhosale, Vaibhavi Lokegaonkar, Vishnu Raj, Gouthaman KV, Sreyan Ghosh, Ramani Duraiswami, Lie Lu, Dinesh Manocha
Published: 9/1/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.CL, cs.CV, cs.LG

arXiv:2608.30125v1 Announce Type: cross Abstract: Current video-to-music (V2M) models lack semantic control and fail to penalize instruction violations, largely due to their reliance on reconstruction objectives and the representational bottleneck of static cross-modal conditioning in Diffusion Auto...

📖 Read original article


287. Balancing Privacy, Utility, and Safety in LLM Alignment through Preference Optimization ​

Author: Dishu Yang, Jingjing Liu, Jize Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2608.30141v1 Announce Type: cross Abstract: Preference optimization is widely used to align large language models with human preferences, but preference-data composition may also influence privacy-relevant memorization. We examine whether adding synthetic privacy-preference pairs to Direct Pre...

📖 Read original article


288. The PUR-1 Cyber-Physical Digital Twin ​

Author: Vasileios Theos, Jonah Lau, Konstantinos Gkouliaras, Zachery Dahm, Konstantinos Vasili, Noah Fillgrove, William Richards, True Miller, Brian Jowers, Stylianos Chatzidakis
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CE, cs.LG

arXiv:2608.30186v1 Announce Type: cross Abstract: Digital twin technologies have the potential to improve operational flexibility and responsiveness capabilities of nuclear systems. To provide decision support, cyber event characterization, state estimation, predictive control, and real-time dynamic...

📖 Read original article


289. Fairness in multi-class multi-group classification problems via contextial coherent risk measures ​

Author: Darinka Dentcheva, Xiangyu Tian
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.30223v1 Announce Type: cross Abstract: We propose a new design of fair classifiers for multi-class classification problems in the presence of vector-valued sensitive attributes. In that scenario each sensitive attribute has multiple values and forms several groups relevant to the fairness...

📖 Read original article


290. Motus2: A Self-Evolving General World Model for Dexterous Manipulation ​

Author: Hongzhe Bi, Zihao Zhou, Yihang Tang, Jingrui Pang, Shuhe Huang, Haitian Liu, Runqing Wang, Shuai Huang, Yichen Wang, Yiming Cheng, Ruowen Zhao, Zhenghua Li, Hengkai Tan, Xiaolong Liu, Jinhui Wan, Jiabao Liu, Min Zhao, Fan Bao, Jun Zhu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG

arXiv:2608.30237v1 Announce Type: cross Abstract: General embodied agents should perceive, predict, act, evaluate, and improve within a unified system. World models have shown great promise in building such agents, yet existing models typically append an action output head to a world simulator, with...

📖 Read original article


291. A Borel Concept Class of VC Dimension One with a Non-PAC Consistent Learner in ZFC ​

Author: Mateus Jesus de Arruda Campos, Gabriel Fernandes, Vinicius de Oliveira Rodrigues
Published: 9/1/2026, 4:00:00 AM
Categories: math.LO, cs.LG, math.PR

arXiv:2608.30246v1 Announce Type: cross Abstract: The fundamental theorem of statistical learning states that, under suitable measurability assumptions, finite Vapnik--Chervonenkis (VC) dimension guarantees that every proper consistent learning rule is probably approximately correct (PAC). Blumer, E...

📖 Read original article


292. Using Prosody to Predict Syntactic Structure ​

Author: Junghyun Min, Alex Warstadt, Tamar I. Regev, Tiago Pimentel, Ethan Gotlieb Wilcox
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.30260v1 Announce Type: cross Abstract: While it is well-established that prosody carries crucial cues for syntactic structure, the degree and nature of correspondence between these two domains remains contested. We investigate the syntax-prosody interface through an information-theoretic ...

📖 Read original article


293. Estimating Population-Risk Curves Along Nonconvex Gradient Flows from the Training Sample ​

Author: Mingzhi Song
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.30261v1 Announce Type: cross Abstract: We estimate the conditional population-risk curve of a realized smooth nonconvex gradient flow from the training sample. Flow approximate leave-one-out (Flow-ALO) propagates a deletion response and evaluates omitted observations at approximate delete...

📖 Read original article


294. Dec-BFTRL: Squre-Root Regret for Decentralized Online Upper-Linearizable Optimization under Separation Access with Application to Continuous Submodular Maximization ​

Author: Yiyang Lu, Mohammad Pedramfar, Vaneet Aggarwal
Published: 9/1/2026, 4:00:00 AM
Categories: math.OC, cs.AI, cs.LG

arXiv:2608.30271v1 Announce Type: cross Abstract: We study decentralized online optimization of upper-linearizable payoffs over an action set under efficient separation access, with applications to online continuous diminishing-return (DR) submodular maximization. We propose Decentralized Barrier Fo...

📖 Read original article


295. Strengthening Recursive Constructions for Zero-Error Shannon Capacity ​

Author: Ravi Tandon
Published: 9/1/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.CO, math.IT

arXiv:2608.30273v1 Announce Type: cross Abstract: The exact Shannon capacity is unknown for every odd cycle beyond the five-cycle $C_5$, making odd cycles a central open problem in zero-error information theory. Improving the known lower bounds requires constructing large independent sets in strong ...

📖 Read original article


296. Beyond Token-Level Guidance: Inference-Time Alignment of Specialized LLMs via Cross-Family Representation Steering ​

Author: Jin Gan, Xin Li, Jun Luo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.30319v1 Announce Type: cross Abstract: Large language models (LLMs) finetuned for specialized domains represent crucial high-impact applications. Inference-time alignment improves safety degraded from specialization finetuning without requiring substantial computational resources, complem...

📖 Read original article


297. Compact and Infinite-Order Error Analysis for Null-Space SVD Estimation ​

Author: Xin Li, Jonathan Cohen, Rami Puzis
Published: 9/1/2026, 4:00:00 AM
Categories: math.ST, cs.LG, stat.TH

arXiv:2608.30374v1 Announce Type: cross Abstract: We study null-space estimation from a noisy matrix. For a simple left null space, we first derive an exact compact expression for the error of the smallest left singular vector. We then give an all-order series for the SVD vector and projector, follo...

📖 Read original article


298. Kathleen Remembers: Length-Invariant One-Shot Recall Without Attention ​

Author: George Fountzoulas
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.30376v1 Announce Type: cross Abstract: Recurrent, attention-free sequence models share a structural weakness: a fading state cannot perform exact recall of something seen once, far in the past. We add to the Kathleen trunk a second memory layer -- a "notebook": a fixed-key holographic (HR...

📖 Read original article


299. Benchmarking External Generalization of SPD Matrix Learning for Resting-State fMRI Connectome Prediction ​

Author: Ce Ju, Antoine Collas, Florent Bouchard, Bertrand Thirion
Published: 9/1/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.30418v1 Announce Type: cross Abstract: Resting-state functional magnetic resonance imaging (rs-fMRI) functional connectivity (FC) matrices are widely used for individual-level prediction, but strong performance within one cohort may not generalize to a new cohort. We ask whether within-da...

📖 Read original article


300. ObjectSplat: Improving Mesh Fidelity and Interactivity for 3D Scenes via Object-Level Mesh Splatting ​

Author: Minhas Kamal, Hiranya Garbha Kumar, Mahedi Kamal, Balakrishnan Prabhakaran
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.30423v1 Announce Type: cross Abstract: Splatting-based algorithms reconstruct photorealistic, real-time-renderable, and mesh-exportable 3D scenes from regular images, but they represent a scene as a single monolithic field. Therefore, the reconstruction has no object-level structure, leav...

📖 Read original article


301. Ceiling-Clipped Acceptance Histograms Indicate Stranded Speed-up in Block-Diffusion Speculative Decoding ​

Author: Ephrem Wu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.30427v1 Announce Type: cross Abstract: Speculative decoding speeds up generation with an efficient draft model (drafter) that proposes tokens for a target model to verify in one pass, preserving the target's output distribution. High-acceptance block-diffusion drafters such as DFlash and ...

📖 Read original article


302. Lies We Can See: Joint Verbal and Non-Verbal Deception by VLM Agents in Embodied Social Interactions ​

Author: Jaewoo Ahn, Junseo Kim, Hyunseo Kim, Heeseung Yun, Jaehyeon Son, Zsolt Kira, Gunhee Kim
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.CV, cs.LG

arXiv:2608.30428v1 Announce Type: cross Abstract: Strategic deception by LLM and VLM agents has emerged as a central AI alignment and safety concern. Social-deduction games (where each player holds a hidden role and communicates with others to deduce identities) serve as the canonical testbed, parti...

📖 Read original article


303. Generalization as a robust performance property of learning-enabled dynamical systems ​

Author: Filippo Fabiani
Published: 9/1/2026, 4:00:00 AM
Categories: eess.SY, cs.LG, cs.SY, math.OC

arXiv:2608.30431v1 Announce Type: cross Abstract: By focusing on algorithmic stability as a means of establishing out-of-sample bounds, we provide a system-theoretic interpretation of generalization in learning-enabled dynamical systems arising in data-driven optimization and feedback control approx...

📖 Read original article


304. Event-Driven Language Models with Sparse Neural Activity for Neuromorphic Hardware ​

Author: Simon Richter, Ruhai Lin, Jason Yik, Taylor Kergan, Rui-Jie Zhu, Farshad Moradi, Jason Eshraghian
Published: 9/1/2026, 4:00:00 AM
Categories: cs.NE, cs.LG

arXiv:2608.30439v1 Announce Type: cross Abstract: Inference with transformer-based large language models (LLMs) is often limited by the memory-bound KV cache and quadratic attention cost. State-space models (SSMs) mitigate this through linear attention and fixed-size recurrent states, but their larg...

📖 Read original article


305. End-to-End Neural Shrinkage of Indefinite Pairwise Correlation Matrices for Small-Cap-Inclusive Portfolios ​

Author: Christian Bongiorno, Lorenzo Villassero
Published: 9/1/2026, 4:00:00 AM
Categories: q-fin.PM, cs.LG, q-fin.ST

arXiv:2608.30446v1 Announce Type: cross Abstract: Small-cap-inclusive equity universes contain recently listed and intermittently traded securities, so enforcing a common look-back discards a substantial fraction of the available information. Pairwise-complete estimation preserves the longest overla...

📖 Read original article


306. VisER: Visual Evidence and Reliance for Object Hallucination Detection in LVLMs ​

Author: Afsaneh Hasanebrahimi, Hanxun Huang, Christopher Leckie, Sarah Erfani
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.30480v1 Announce Type: cross Abstract: Object hallucination remains a persistent reliability issue in large vision-language models, where generated object mentions may sound plausible but lack visual grounding. Recent training-free detectors use internal signals such as token likelihood, ...

📖 Read original article


307. Two Centuries of Sexism in British Parliament: A Computational Analysis of Women's Representation in the Hansard Corpus ​

Author: Mohammad Omar Khursheed, Mandira Sawkar, Ashiqur R. KhudaBukhsh
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.CY, cs.LG

arXiv:2608.30485v1 Announce Type: cross Abstract: The language a legislature uses to debate women's rights, even in favour of them, encodes systematic patterns of sexism that persist across two centuries. In this work, we analyse 6,531 speeches over 200 years of UK parliamentary debate (Hansard, 180...

📖 Read original article


308. TSExplorer: An interactive data annotation and exploration tool for time-series data ​

Author: Einari Vaaras, Manu Airaksinen, Okko R"as"anen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.HC, cs.AI, cs.LG, cs.SE

arXiv:2608.30514v1 Announce Type: cross Abstract: We present TSExplorer, a cross-platform tool for interactive annotation and exploration of time-series data. The tool enables users to inspect high-dimensional datasets through multiple complementary 2D visualizations derived from high-dimensional fe...

📖 Read original article


309. Beamforming Design Via GNN in mmWave Cell-Free Massive MIMO Using Sub-6 GHz CSI ​

Author: Sina Tavakolian, Abolfazl Zakeri, Ahmed Alkhateeb, Markku Juntti, Nhan Thanh Nguyen
Published: 9/1/2026, 4:00:00 AM
Categories: eess.SP, cs.LG

arXiv:2608.30524v1 Announce Type: cross Abstract: Beamforming methods in millimeter-wave (mmWave) cell-free massive multiple-input multiple-output (CFmMIMO) systems require accurate channel state information (CSI), whose acquisition entails significant training overhead. This paper shows that fully ...

📖 Read original article


310. Minerals in the Wild: A Hyperspectral-XRF Dataset for Elemental Composition Estimation ​

Author: Eleftheria Tetoula-Tsonga (Institute of Communication and Computer Systems, Athens, Greece), George Arvanitakis (Geonova, Athens, Greece), Theodoros Giannakas (Institute of Communication and Computer Systems, Athens, Greece)
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.30537v1 Announce Type: cross Abstract: Rapid mineral characterization is essential for applications ranging from mineral exploration to industrial ore processing. To this end, Hyperspectral Imaging (HSI) has emerged as a promising sensing modality thanks to its fine spectral resolution, e...

📖 Read original article


311. Informative Label Missingness in Multiclass Classification Information Geometry and Excess Risk ​

Author: Fariborz Setoudehtazang, Geoffrey J. McLachlan
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.30561v1 Announce Type: cross Abstract: Informative label missingness can change the usual efficiency ordering between completely and partially labelled classifiers because the pattern of missing labels may itself carry information about the classification model. We develop a general likel...

📖 Read original article


312. Automated Testing of LLM-Based Post Hoc Explainers Using Model Checking as an Oracle ​

Author: Dennis Gross, Helge Spieker
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.30581v1 Announce Type: cross Abstract: Large language models (LLMs) are used as post hoc explainers of sequential decision-making policies, producing natural-language explanations of why an action was chosen. However, LLMs often generate plausible but incorrect statements, and no existing...

📖 Read original article


313. Reading the News: Adapting Large Language Models to Swedish Journalism Through Continued Pre-Training ​

Author: Lukas Borggren, Jenny Kunz, Marco Kuhlmann
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.30609v1 Announce Type: cross Abstract: Large language models are increasingly capable in general, but their utility can remain modest in niche or understudied areas. One approach to address this limitation is to specialise existing models through additional training on target-domain corpo...

📖 Read original article


314. GMTS: Gradient Magnitude-based Token Selection Improves RLVR Training for LLM Reasoning ​

Author: Outongyi Lv, Yuanwei Zhang, Xiaoqun Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.30632v1 Announce Type: cross Abstract: Reinforcement learning (RL), particularly RL with Verifiable Rewards (RLVR), has recently emerged as a central paradigm for enhancing large language models' (LLMs) reasoning abilities, demonstrating remarkable effectiveness across reasoning tasks. Re...

📖 Read original article


315. Quantum-Grassmann-Plucker Token Mixing for Deep Learning-Based Post-Disaster Damage Assessment ​

Author: Kooroush Farahkhah, Umut Lagap, Taha Rezaei, Saman Ghaffarian
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.30633v1 Announce Type: cross Abstract: Timely post-disaster building damage assessment from satellite imagery is a critical engineering decision support task, yet it remains constrained by class imbalance, ambiguous intermediate damage states, and limited cross-event transferability. This...

📖 Read original article


316. BiG-SURE - Bipartite Graph for Semantic Uncertainty and Reliability Estimation of LLMs ​

Author: Debarpan Bhattacharya, Malay Phadke, Sriram Ganapathy
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.30646v2 Announce Type: cross Abstract: Reliable uncertainty estimation is a crucial requirement for deploying large language models (LLMs) and vision-language models (VLMs) in safety-critical settings, especially when the model parameters are not accessible (black-box). We propose BiG-SUR...

📖 Read original article


317. What It Costs to Compose, Rebuild, and Correct Precomputed Memory ​

Author: Asa Shepard
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.30647v1 Announce Type: cross Abstract: Language models can answer from precomputed memory, a model's saved reading of a body of material, reused across requests instead of read again at each. This paper maps where that practice preserves correctness and the conditions under which it fails...

📖 Read original article


318. Fine-Grained Multi Image Object Hallucination Benchmark ​

Author: Joonki Min, Chaeyun Kim, Hyungwook Choi, Yejin Kim, Kihyun Kim, Yohan Jo, Joonseok Lee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.30653v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are increasingly deployed in multi-image scenarios requiring complex reasoning across visual contexts. However, current MLLMs remain fundamentally limited by object hallucination-generating plausible yet factu...

📖 Read original article


319. An Agentic Retrobiosynthesis Framework with Learned Frontier Selection ​

Author: Philippe Meyer, Guillaume Gricourt, Thomas Duigou, Joan H'erisson, Jean-Loup Faulon
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.30702v1 Announce Type: cross Abstract: Large language models are increasingly used as agents for multistep retrosynthesis, raising the question of how much their search policy contributes independently of the underlying reaction model. We investigate this question in a biological setting ...

📖 Read original article


320. SingProbe Technical Report ​

Author: Sing Team
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.CL, cs.LG

arXiv:2608.30703v1 Announce Type: cross Abstract: Runtime guardrails are essential for reliable large language model (LLM) deployment, yet existing approaches typically rely on independent, external models that introduce additional inference cost, delayed safety signals, and a capacity mismatch with...

📖 Read original article


321. Conjoint Audio-to-Spikes Encoding and Processing for Efficient Neuromorphic Speech Recognition ​

Author: Valentin M. Meunier, Am'elie Gruel, Pierre Lewden, Adrien F. Vincent, Sylvain Sa"ighi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.NE, cs.AI, cs.LG, eess.AS

arXiv:2608.30792v1 Announce Type: cross Abstract: Obtaining data from neuromorphic sensors and processing it with Spiking Neural Networks is a promising solution to lower the energy cost of artificial intelligence. The current rarity of natively neuromorphic datasets promotes the development of soft...

📖 Read original article


322. Uncertainty-Aware End-to-End AI Weather Forecasting: Disentangling Observation and Model Contributions ​

Author: Rodrigo Almeida, Noelia Otero, Jost Arndt, Simon Baur, Wojciech Samek, Jackie Ma
Published: 9/1/2026, 4:00:00 AM
Categories: physics.ao-ph, cs.LG

arXiv:2608.30795v1 Announce Type: cross Abstract: End-to-end weather forecasting systems produce skillful global gridded and station forecasts directly from raw Earth observations, replacing the numerical weather prediction pipeline, including data assimilation, at a fraction of its cost. These syst...

📖 Read original article


323. TopoCompress: Long Context Compression via Graph-Wired Semantic Trajectories ​

Author: Daniel Agyei Asante, Yang Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.30811v2 Announce Type: cross Abstract: Long-context compression is essential for reducing the cost and latency of large language model inference. However, existing methods can fragment important evidence, require additional training or alignment, and often depend on the target model for e...

📖 Read original article


324. Linguistic Distance Segregates Latent Representations in Automatic Speech Recognition Systems ​

Author: Ting-Hui Cheng, Line Katrine Harder Clemmensen, Sneha Das
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.30853v1 Announce Type: cross Abstract: While automatic speech recognition (ASR) models have achieved remarkable improvements in recent years, performance disparities persist across different speaker populations. One such disparity is for speakers whose first languages (L1) are from famili...

📖 Read original article


325. Safety Screening for Voltage Control in Active Distribution Grids via Distributionally Robust Conformal Screening ​

Author: Sarra Bouchkati, Petros Ellinas, Adriana Geisler, Steffen Kortmann, Johanna Vorwerk, Spyros Chatzivasiliadis, Andreas Ulbig
Published: 9/1/2026, 4:00:00 AM
Categories: eess.SY, cs.AI, cs.LG, cs.SY

arXiv:2608.30889v1 Announce Type: cross Abstract: Deploying a new control policy for voltage control in active distribution grids requires evidence that physical limits will be satisfied before the policy is tested on the physical grid. This assessment is difficult for two reasons. First, simulation...

📖 Read original article


326. CoJEPA: Combining Contrastive Learning and JEPA for Global-Local Music Representations ​

Author: Gabriel Meseguer-Brocal, Yuexuan Kong, Romain Hennequin
Published: 9/1/2026, 4:00:00 AM
Categories: cs.SD, cs.AI, cs.LG, eess.AS, eess.SP

arXiv:2608.30974v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architecture (JEPA) has shown strong performance in learning rich representations through self-supervised prediction in latent space. However, it typically relies on teacher--student architecture with an EMA to stabilise tr...

📖 Read original article


327. Stick to What You Know: A Study of Knowledge-Aligned Supervised Fine-Tuning ​

Author: Arthur Becker, Jakob Kemmler, David Thulke, Christine Sch"afer, Christian Dugast, Hermann Ney
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2608.30987v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) trains a base language model to imitate target responses, and these targets may require knowledge the base model has not robustly internalized. We study this as a source of hallucinations and frame a group of mitigation m...

📖 Read original article


328. Learning the Geometry of Admissible Hypotheses through Inductive Bias in Training Distributions ​

Author: James Crowley, Faez Ahmed, Anton van Beek
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.31028v1 Announce Type: cross Abstract: Scientific discovery often requires reasoning over competing hypotheses that are consistent with experimental observations. For mixed-variable and combinatorial hypothesis spaces, however, constructing probabilistic representations remains challengin...

📖 Read original article


329. Driving on Memory ​

Author: Christian L"owens, Thorben Funke, Alexandru Paul Condurache
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, cs.RO

arXiv:2608.31029v1 Announce Type: cross Abstract: End-to-end autonomous driving models plan future trajectories from raw sensor input. While earlier driving benchmarks often measured deviation from the human trajectory, current benchmarks such as NAVSIM and Bench2Drive evaluate models with richer si...

📖 Read original article


330. Segmentation of Bovid Dentition Under Imperfect Annotations: A Comparative Study of Convolutional and Attention Models ​

Author: Keith G. Mills, Evan B. Sanders, Gregory J. Matthews, Juliet K. Brophy
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.31052v1 Announce Type: cross Abstract: Semantic segmentation decomposes an image into distinct mask regions corresponding to different object categories, such as people, cars, signs or buildings. Advances in machine learning (ML) have shifted this task away from traditional rule-based heu...

📖 Read original article


331. Learning to Evaluate Before Improving: Automatic Rubric Induction for Automatic Research Agents ​

Author: Xuehai Wang, Haowei Qin, Tongxin Liu, Junkai Li, Buqiang Xu, Jintian Zhang, Yijun Chen, Zirui Xue, Shumin Deng
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.IR, cs.LG, cs.MA, cs.SE

arXiv:2608.31076v1 Announce Type: cross Abstract: Autonomous scientific research agents are increasingly applied to end-to-end scientific workflows, including literature review, data analysis, experimentation, and report generation. However, open-ended research tasks often do not clearly specify the...

📖 Read original article


332. Minimax bounds for watermarked and masked recursive discrete distribution estimation ​

Author: Millen Kanabar, Michael Gastpar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.IT, cs.LG, math.IT, math.ST, stat.TH

arXiv:2608.31091v1 Announce Type: cross Abstract: Watermarking has been proposed as a way to identify synthetic samples in estimation settings where no metadata is available to distinguish them from real samples, but its precise effects remain unexplored. In the absence of a distinguishing mechanism...

📖 Read original article


333. One Adapter, Many Tasks: Task-Conditioned Feature Transformations for Continual Learning ​

Author: Yunxiang Fu, Meng Lou, Yizhou Yu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2608.31096v1 Announce Type: cross Abstract: Class-incremental learning (CIL) requires a model to incrementally learn tasks that contain new classes without accessing earlier training data while preserving the ability to recognize all seen classes. Recently, pretrained-model-based approaches ha...

📖 Read original article


334. LLM Post-Training as Brownfield Maintenance: An Industrial Perspective on Dataware Engineering ​

Author: Gopi Krishnan Rajbahadur, Amir M. Ebrahimi, Boyuan Chen, Ahmed E. Hassan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.SE, cs.AI, cs.LG

arXiv:2608.31102v1 Announce Type: cross Abstract: Industrial post-training is a brownfield regime. Teams inherit a deployed checkpoint and must land targeted improvements under fixed compute and mixture budgets without regressing the rest. The maintained artifact is increasingly dataware: behavior g...

📖 Read original article


335. "Train classical, deploy quantum" requires rethinking generalization ​

Author: Snehal Raj, Natansh Mathur, Alejandro Perdomo-Ortiz
Published: 9/1/2026, 4:00:00 AM
Categories: quant-ph, cs.LG

arXiv:2608.31117v1 Announce Type: cross Abstract: Generative models have become central across science and industry, from image and text synthesis to the design of molecules and materials. Quantum generative models are considered one of the most promising applications for quantum computers, since a ...

📖 Read original article


336. Implementing neural network mixed-effects models in Template Model Builder (TMB) ​

Author: Nan Zheng, Hoi Yiu Cheung, Vibhu Sharma, James T. Thorson, Noel G. Cadigan
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2608.31133v1 Announce Type: cross Abstract: Neural network mixed-effects models (NMMs) have gained traction by combining the strong representation and predictive power of artificial neural networks with the capacity of mixed-effects modeling to capture complex correlation structures. However, ...

📖 Read original article


337. GFlowNets and variational inference ​

Author: Esmeralda S. Whitammer, Salem Lahlou, Tristan Deleu, Xu Ji, Edward Hu, Katie Everett, Dinghuai Zhang, Yoshua Bengio
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2210.00580v4 Announce Type: replace Abstract: This paper builds bridges between two families of probabilistic algorithms: (hierarchical) variational inference (VI), which is typically used to model distributions over continuous spaces, and generative flow networks (GFlowNets), which have been ...

📖 Read original article


338. Delta-AI: Local objectives for amortized inference in sparse graphical models ​

Author: Jean-Pierre Falet, Hae Beom Lee, Esmeralda S. Whitammer, Chen Sun, Dragos Secrieru, Thomas Jiralerspong, Dinghuai Zhang, Guillaume Lajoie, Yoshua Bengio
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2310.02423v3 Announce Type: replace Abstract: We present a new algorithm for amortized inference in sparse probabilistic graphical models (PGMs), which we call $\Delta$-amortized inference ($\Delta$-AI). Our approach is based on the observation that when the sampling of variables in a PGM is s...

📖 Read original article


339. Expected flow networks in stochastic environments and two-player zero-sum games ​

Author: Marco Jiralerspong, Bilun Sun, Danilo Vucetic, Tianyu Zhang, Yoshua Bengio, Gauthier Gidel, Esmeralda S. Whitammer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.GT

arXiv:2310.02779v3 Announce Type: replace Abstract: Generative flow networks (GFlowNets) are sequential sampling models trained to match a given distribution. GFlowNets have been successfully applied to various structured object generation tasks, sampling a diverse set of high-reward objects quickly...

📖 Read original article


340. Amortizing intractable inference in large language models ​

Author: Edward J. Hu, Moksh Jain, Eric Elmoznino, Younesse Kaddar, Guillaume Lajoie, Yoshua Bengio, Esmeralda S. Whitammer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2310.04363v3 Announce Type: replace Abstract: Autoregressive large language models (LLMs) compress knowledge from their training data through next-token conditional distributions. This limits tractable querying of this knowledge to start-to-end autoregressive sampling. However, many tasks of i...

📖 Read original article


341. SUB-PLAY: Adversarial Policies against Partially Observed Multi-Agent Reinforcement Learning Systems ​

Author: Oubo Ma, Yuwen Pu, Linkang Du, Yang Dai, Ruo Wang, Xiaolei Liu, Yingcai Wu, Shouling Ji
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2402.03741v4 Announce Type: replace Abstract: Recent advancements in multi-agent reinforcement learning (MARL) have opened up vast application prospects, such as swarm control of drones, collaborative manipulation by robotic arms, and multi-target encirclement. However, potential security thre...

📖 Read original article


342. Branch Scaling Manifests as Implicit Architectural Regularization for Improving Generalization in Overparameterized ResNets ​

Author: Zixiong Yu, Guhan Chen, Jianfa Lai, Bohan Li, Songtao Tian
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, math.ST, stat.TH

arXiv:2403.04545v4 Announce Type: replace Abstract: Scaling factors in residual branches have emerged as a prevalent method for boosting neural network performance, especially in normalization-free architectures. While prior work has primarily examined scaling effects from an optimization perspectiv...

📖 Read original article


343. Understanding Deep Learning via Notions of Rank ​

Author: Noam Razin
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.NE, stat.ML

arXiv:2408.02111v4 Announce Type: replace Abstract: Despite the extreme popularity of deep learning in science and industry, its formal understanding is limited. This thesis puts forth notions of rank as key for developing a theory of deep learning, focusing on the fundamental aspects of generalizat...

📖 Read original article


344. Adaptive teachers for amortized samplers ​

Author: Minsu Kim, Sanghyeok Choi, Taeyoung Yun, Emmanuel Bengio, Leo Feng, Jarrid Rector-Brooks, Sungsoo Ahn, Jinkyoo Park, Esmeralda S. Whitammer, Yoshua Bengio
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2410.01432v3 Announce Type: replace Abstract: Amortized inference is the task of training a parametric model, such as a neural network, to approximate a distribution with a given unnormalized density where exact sampling is intractable. When sampling is implemented as a sequential decision-mak...

📖 Read original article


345. MEGA: Message Passing Neural Networks for Multigraphs with EdGe Attributes ​

Author: H. \c{C}a\u{g}r{\i} Bilgi, Kubilay Atasu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2412.00241v3 Announce Type: replace Abstract: Edge-attributed multigraphs, in which multiple edges with distinct attributes connect the same pair of nodes, arise naturally in many real-world systems. In these graphs, effective learning requires preserving information from repeated interactions...

📖 Read original article


346. H-FedSN: Personalized Sparse Networks for Efficient and Accurate Hierarchical Federated Learning for IoT Applications ​

Author: Jiechao Gao, Yuangang Li, Jie Wang, Yue Zhao, Michael Lepech, Brad Campbell
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2412.06210v3 Announce Type: replace Abstract: With the rapid development of the Internet of Things (IoT), federated learning (FL) has gained increasing attention for its privacy-preserving use of distributed data. However, conventional two-tier FL architectures are poorly suited to the hierarc...

📖 Read original article


347. Alert: Learning Trigger Functions for Early Classification of Time Series using Deep-RL ​

Author: Aur'elien Renault, Alexis Bondu, Antoine Cornu'ejols, Vincent Lemaire
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2502.06584v2 Announce Type: replace Abstract: Early Classification of Time Series (ECTS) is vital in fields like industrial monitoring and medical triage, where quick and accurate predictions are essential. One of the core challenges lies in the trigger function, which decides when to make a p...

📖 Read original article


348. You Do Not Fully Utilize Transformer's Representation Capacity ​

Author: Gleb Gerasimov, Yaroslav Aksenov, Nikita Balagansky, Viacheslav Sinii, Daniil Gavrilov
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2502.09245v3 Announce Type: replace Abstract: In contrast to RNNs, which compress their history into a single hidden state, Transformers can attend to all past tokens directly. However, standard Transformers rely solely on the hidden state from the previous layer to represent the entire contex...

📖 Read original article


349. RSPO: Regularized Self-Play Alignment of Large Language Models ​

Author: Xiaohang Tang, Sangwoong Yoon, Seongho Son, Huizhuo Yuan, Quanquan Gu, Ilija Bogunovic
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2503.00030v3 Announce Type: replace Abstract: Self-play-based policy optimization has emerged as an effective approach for fine-tuning large language models (LLMs), formulating preference optimization as a two-player game. However, the regularization with respect to the reference policy, which...

📖 Read original article


350. Temporal Analysis of NetFlow Datasets for Network Intrusion Detection Systems ​

Author: Majed Luay, Siamak Layeghy, Seyedehfaezeh Hosseininoorbin, Mohanad Sarhan, Nour Moustafa, Marius Portmann
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CR, cs.NI

arXiv:2503.04404v3 Announce Type: replace Abstract: This paper investigates the temporal analysis of NetFlow datasets for machine learning (ML)-based network intrusion detection systems (NIDS). Although many previous studies have highlighted the critical role of temporal features, such as inter-pack...

📖 Read original article


351. Semantics at an Angle: When Cosine Similarity Works Until It Doesn't ​

Author: Kisung You
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2504.16318v4 Announce Type: replace Abstract: Cosine similarity is a standard comparison rule for learned representations in information retrieval, natural language processing, computer vision, and multimodal learning. Its popularity is well founded: it removes positive radial scale, is comput...

📖 Read original article


352. Graph Representational Learning: When Does More Expressivity Hurt Generalization? ​

Author: Sohir Maskey, Raffaele Paolino, Fabian Jogl, Gitta Kutyniok, Johannes F. Lutzeyer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2505.11298v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are powerful tools for learning on structured data, yet the relationship between their expressivity and predictive performance remains unclear. We introduce a family of premetrics that capture different degrees of struc...

📖 Read original article


353. Large-Scale Bayesian Tensor Reconstruction via Approximate Message Passing ​

Author: Bingyang Cheng, Zhongtao Chen, Yichen Jin, Hao Zhang, Chen Zhang, Edmund Y. Lam, Yik-Chung Wu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, eess.SP

arXiv:2505.16305v3 Announce Type: replace Abstract: While CANDECOMP/PARAFAC (CP) decomposition (CPD) is fundamental for tensor reconstruction, Bayesian CPD often scales poorly because variational updates require repeated matrix inversions. We develop CP generalized approximate message passing (CP-GA...

📖 Read original article


354. Watch your steps: Dormant Adversarial Behaviors that Activate upon LLM Finetuning ​

Author: Thibaud Gloaguen, Mark Vero, Robin Staab, Martin Vechev
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CR

arXiv:2505.16567v4 Announce Type: replace Abstract: Finetuning open-weight Large Language Models (LLMs) is standard practice for achieving task-specific performance improvements. Until now, finetuning has been regarded as a controlled and secure process in which training on benign datasets leads to ...

📖 Read original article


355. Kronecker Factorization Improves Efficiency and Interpretability of Sparse Autoencoders ​

Author: Vadim Kurochkin, Yaroslav Aksenov, Daniil Laptev, Daniil Gavrilov, Nikita Balagansky
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2505.22255v4 Announce Type: replace Abstract: Sparse Autoencoders (SAEs) decompose language-model activations into sparse, interpretable features, but standard encoders usually treat the latent dictionary as a flat set of independent coordinates, leaving hierarchy and feature interactions to e...

📖 Read original article


356. EquiReg: Equivariance Regularized Diffusion for Inverse Problems ​

Author: Bahareh Tolooshams, Aditi Chandrashekar, Rayhan Zirvi, Abbas Mammadov, Jiachen Yao, Chuwei Wang, Anima Anandkumar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2505.22973v3 Announce Type: replace Abstract: Diffusion models represent the state-of-the-art for solving inverse problems such as image restoration tasks. Diffusion-based inverse solvers incorporate a likelihood term to guide prior sampling, generating data consistent with the posterior distr...

📖 Read original article


357. Federated Learning for MRI-based BrainAGE: a multicenter study on post-stroke functional outcome prediction ​

Author: Vincent Roca, Marc Tommasi, Paul Andrey, Aur'elien Bellet, Markus D. Schirmer, Hilde Henon, Laurent Puy, Julien Ramon, Gr'egory Kuchcinski, Martin Bretzner, Renaud Lopes
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC

arXiv:2506.15626v3 Announce Type: replace Abstract: $\textbf{Objective:}$ Brain-predicted age difference (BrainAGE) is a neuroimaging biomarker reflecting brain health. However, training robust BrainAGE models requires large datasets, often restricted by privacy concerns. This study evaluates the pe...

📖 Read original article


358. Discrete Compositional Generation via General Soft Operators and Robust Reinforcement Learning ​

Author: Marco Jiralerspong, Esther Derman, Danilo Vucetic, Esmeralda S. Whitammer, Bilun Sun, Tianyu Zhang, Pierre-Luc Bacon, Gauthier Gidel
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2506.17007v3 Announce Type: replace Abstract: A major bottleneck in scientific discovery consists of narrowing an exponentially large set of objects, such as proteins or molecules, to a small set of promising candidates with desirable properties. While this process can rely on expert knowledge...

📖 Read original article


359. Training-free LLM Verification via Recycling Few-shot Examples ​

Author: Dongseok Lee, Jimyung Hong, Dongyoung Kim, Jaehyung Kim
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2506.17251v3 Announce Type: replace Abstract: Although large language models (LLMs) have achieved remarkable performance, the inherent stochasticity of their reasoning processes and varying conclusions present significant challenges. Majority voting or Best-of-N with external verifiers has bee...

📖 Read original article


360. A foundation model with multi-variate parallel attention to generate neuronal activity ​

Author: Francesco Carzaniga, Michael Hersche, Abu Sebastian, Kaspar Schindler, Abbas Rahimi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2506.20354v3 Announce Type: replace Abstract: Learning from multi-variate time-series with heterogeneous channel configurations remains a fundamental challenge for deep neural networks, particularly in clinical domains such as intracranial electroencephalography (iEEG), where channel setups va...

📖 Read original article


361. A Cycle-Consistency Constrained Framework for Dynamic Solution Space Reduction in Noninjective Regression ​

Author: Hanzhang Jia, Yi Gao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2507.04659v3 Announce Type: replace Abstract: To address the challenges posed by the heavy reliance of multi-output models on preset probability distributions and embedded prior knowledge in non-injective regression tasks, this paper proposes a cycle consistency-based data-driven training fram...

📖 Read original article


362. A Conditional GAN for Tabular Data Generation with Probabilistic Sampling of Latent Subspaces ​

Author: Leonidas Akritidis, Panayiotis Bozanis
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2508.00472v3 Announce Type: replace Abstract: The tabular form constitutes the standard way of representing data in relational database systems and spreadsheets. But, similarly to other forms, tabular data suffers from class imbalance, a problem that causes serious performance degradation in a...

📖 Read original article


363. StructSynth: Dependency Graphs as Generation Plans for Low-Data Tabular Synthesis with Language Models ​

Author: Siyi Liu, Yujia Zheng, Haoyang Li, Yongqi Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2508.02601v2 Announce Type: replace Abstract: Tabular data derives its value from inter-feature dependencies, yet preserving them during synthesis is fragile when samples are scarce. Existing approaches either learn dependencies implicitly through distribution fitting, rely on statistical grap...

📖 Read original article


364. Agnostics: Learning to Code in Any Programming Language via Reinforcement with a Universal Learning Environment ​

Author: Aleksander Boruch-Gruszecki, Yangtian Zi, Zixuan Wu, Tejas Oberoi, Carolyn Jane Anderson, Joydeep Biswas, Arjun Guha
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.PL

arXiv:2508.04865v4 Announce Type: replace Abstract: Large language models (LLMs) already excel at writing code in high-resource languages such as Python and JavaScript, yet stumble on low-resource languages that remain essential to science and engineering. Besides the obvious shortage of pre-trainin...

📖 Read original article


365. Integrating attention into explanation frameworks for language and vision transformers ​

Author: Marte Eggen, Jacob Lysn{\ae}s-Larsen, Inga Str"umke
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2508.08966v2 Announce Type: replace Abstract: The attention mechanism lies at the core of the transformer architecture, providing an interpretable model-internal signal that has motivated a growing interest in attention-based model explanations. Although attention weights do not directly deter...

📖 Read original article


366. EEGDM: Learning EEG Representation with Latent Diffusion Model ​

Author: Shaocong Wang, Tong Liu, Yihan Li, Ming Li, Kairui Wen, Pei Yang, Wenqi Ji, Minjing Yu, Yong-Jin Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2508.20705v4 Announce Type: replace Abstract: Recent advances in self-supervised learning for EEG representation have largely relied on masked reconstruction, where models are trained to recover randomly masked signal segments. While effective at modeling local dependencies, the training objec...

📖 Read original article


367. SHAKE-GNN: Scalable Hierarchical Kirchhoff-Forest Graph Neural Network ​

Author: Zhipu Cui, Johannes Lutzeyer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2509.22100v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have achieved remarkable success across a range of learning tasks. However, scaling GNNs to large graphs remains a significant challenge, especially for graph-level tasks. In this work, we introduce SHAKE-GNN, a novel s...

📖 Read original article


368. Data-to-Energy Stochastic Dynamics ​

Author: Kirill Tamogashev, Esmeralda S. Whitammer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2509.26364v3 Announce Type: replace Abstract: The Schr"odinger bridge problem is concerned with finding a stochastic dynamical system bridging two marginal distributions that minimises a certain transportation cost. This problem, which represents a generalisation of optimal transport to the s...

📖 Read original article


369. Multi-Marginal Flow Matching with Adversarially Learnt Interpolants ​

Author: Oskar Kviman, Kirill Tamogashev, Nicola Branchini, V'ictor Elvira, Jens Lagergren, Esmeralda S. Whitammer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.01159v3 Announce Type: replace Abstract: Learning the dynamics of a process given sampled observations at several time points is an important but difficult task in many scientific applications. When no ground-truth trajectories are available, but one has only snapshots of data taken at di...

📖 Read original article


370. MolGA: Molecular Graph Adaptation with Pre-trained 2D Graph Encoder ​

Author: Xingtong Yu, Chang Zhou, Xinming Zhang, Yuan Fang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2510.07289v3 Announce Type: replace Abstract: Molecular graph representation learning is widely used in chemical and biomedical research. While pre-trained 2D graph encoders have demonstrated strong performance, they overlook the rich molecular domain knowledge associated with submolecular ins...

📖 Read original article


371. Personalized Treatment Outcome Prediction from Scarce Data via Dual-Channel Knowledge Distillation and Adaptive Fusion ​

Author: Wenjie Chen, Li Zhuang, Ziying Luo, Yu Liu, Jiahao Wu, Shengcai Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2510.26444v2 Announce Type: replace Abstract: Personalized treatment outcome prediction based on trial data for small-sample and rare patient groups is a critical task in precision medicine. However, the high cost and scarcity of trial data limit the prediction performance. To address this iss...

📖 Read original article


372. Structure-Preserving Physics-Informed Neural Network for the Korteweg--de Vries (KdV) Equation ​

Author: Victory Obieke, Emmanuel Oguadimma
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, math-ph, math.MP, nlin.PS, physics.flu-dyn

arXiv:2511.00418v3 Announce Type: replace Abstract: Physics-Informed Neural Networks (PINNs) offer a flexible framework for solving nonlinear partial differential equations (PDEs), yet conventional implementations often fail to preserve key physical invariants during long-term integration. This pape...

📖 Read original article


373. Deep Reinforcement Learning for Dynamic Origin-Destination Matrix Estimation in Microscopic Traffic Simulations Considering Credit Assignment ​

Author: Donggyu Min, Seongjin Choi, Dong-Kyu Kim
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.06229v4 Announce Type: replace Abstract: This paper focuses on dynamic origin-destination matrix estimation (DODE), a crucial calibration process necessary for the effective application of microscopic traffic simulations. The fundamental challenge of the DODE problem in microscopic simula...

📖 Read original article


374. Uncertainty Makes It Stable: Curiosity-Driven Quantized Mixture-of-Experts ​

Author: Sebasti'an Andr'es Cajas Ord'o~nez, Luis Fernando Torres Torres, Mackenzie J. Meni, Carlos Andr'es Duran Paredes, Eric Arazo, Cristian Bosch, Ricardo Simon Carbajo, Yuan Lai, Leo Anthony Celi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2511.11743v4 Announce Type: replace Abstract: Deploying deep neural networks on resource-constrained devices faces two critical challenges: maintaining accuracy under aggressive quantization while ensuring predictable inference latency. We present a curiosity-driven quantized Mixture-of-Expert...

📖 Read original article


375. The Double-Edged Nature of the Rashomon Set for Trustworthy Machine Learning ​

Author: Ethan Hsu, Harry Chen, Chudi Zhong, Lesia Semenova
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2511.21799v2 Announce Type: replace Abstract: Real-world machine learning (ML) pipelines rarely produce a single model; instead, they produce a Rashomon set of many near-optimal ones. We show that this multiplicity reshapes key aspects of trustworthiness. At the individual-model level, sparse ...

📖 Read original article


376. SpecPV: Improving Self-Speculative Decoding for Long-Context Generation via Partial Verification ​

Author: Zhendong Tan, Xingjun Zhang, Chaoyi Hu, Junjie Peng, Kun Xia
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2512.02337v2 Announce Type: replace Abstract: Growing demands from tasks like code generation, deep reasoning, and long-document understanding have made long-context generation a crucial capability for large language models (LLMs). Speculative decoding is one of the most direct and effective a...

📖 Read original article


377. ScalePRM: Training Process Reward Models by Scaling Verification Compute Without Ground Truth ​

Author: Salman Rahman, Sruthi Gorantla, Arpit Gupta, Swastik Roy, Nanyun Peng, Yang Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2512.03244v2 Announce Type: replace Abstract: Training process reward models (PRMs) requires step-level correctness labels, obtained either through expensive human annotation or by relying on ground-truth answers, limiting the ability to scale process-level supervision. We propose ScalePRM, wh...

📖 Read original article


378. Better World Models Can Lead to Better Post-Training Performance ​

Author: Prakhar Gupta, Henry Conklin, Sarah-Jane Leslie, Andrew Lee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2512.03400v2 Announce Type: replace Abstract: We study how explicit world-modeling objectives affect the internal representations and downstream capability of Transformers, using Rubik's Cubes as our training domain. We ask: (1) how does explicitly pretraining a world model affect a model's la...

📖 Read original article


379. Kascade: A Practical Sparse Attention Method for Long-Context LLM Inference ​

Author: Dhruv Deshmukh, Saurabh Goyal, Nipun Kwatra, Ramachandran Ramjee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC

arXiv:2512.16391v2 Announce Type: replace Abstract: Attention is the dominant source of latency during long-context LLM inference, an increasingly popular workload with reasoning models and RAG. We propose Kascade, a training-free sparse attention method that leverages known observations such as 1) ...

📖 Read original article


380. KV Admission: Learning What to Write for Efficient Long-Context LLM Inference ​

Author: Yen-Chieh Huang, Pi-Cheng Hsiu, Rui Fang, Ming-Syan Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2512.17452v4 Announce Type: replace Abstract: Long-context LLM inference is bottlenecked by the quadratic attention complexity and linear Key-Value (KV) cache growth. Prior approaches mitigate this via post-hoc selection or eviction but overlook the root inefficiency: indiscriminate token admi...

📖 Read original article


381. Audit Me If You Can: Query-Efficient Active Fairness Auditing of Black-Box LLMs ​

Author: David Hartmann, Lena Pohlmann, Lelia Hanslik, Noah Gie{\ss}ing, Bettina Berendt, Pieter Delobelle
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CY

arXiv:2601.03087v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit systematic biases across demographic groups. Auditing is proposed as an accountability tool for black-box LLM applications, but suffers from resource-intensive query access. We conceptualise auditing as uncertai...

📖 Read original article


382. Federated Personalization of Early-Exit Networks ​

Author: Boyi Liu, Zimu Zhou, Cheng Fang, Yongxin Tong
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2601.10015v2 Announce Type: replace Abstract: Personalized Federated Learning (PFL) excels at tailoring client-specific models, which is particularly critical for decentralized and heterogeneous data environments, yet existing methods produce static models with a fixed tradeoff between accurac...

📖 Read original article


383. Mechanism Shift During Post-training from Autoregressive to Masked Diffusion Language Models ​

Author: Injin Kong, Hyoungjoon Lee, Yohan Jo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2601.14758v5 Announce Type: replace Abstract: Post-training pretrained autoregressive models (ARMs) into masked diffusion models (MDMs) provides an efficient route to diffusion language modeling, but it remains unclear whether the resulting models reuse inherited autoregressive computation or ...

📖 Read original article


384. Securing Time Integrity in Energy IoT Against Clock Drift and Y2K38 Failures ​

Author: Saeid Jamshidi, Foutse Khomh, Carol Fung, Omar Abdul Wahab, Rolando Herrero
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2601.23147v3 Announce Type: replace Abstract: Time integrity across distributed Internet of Things (IoT) devices is fundamental to reliable sensing, control, and security in energy cyber-physical systems. However, operational energy IoT systems remain vulnerable to clock-drift escalation, time...

📖 Read original article


385. Semi-supervised CAPP Transformer Learning via Pseudo-labeling ​

Author: Dennis Gross, Helge Spieker, Arnaud Gotlieb, Emmanuel Stathatos, Panorios Benardos, George-Christopher Vosniakos
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.01419v2 Announce Type: replace Abstract: High-level Computer-Aided Process Planning (CAPP) generates manufacturing process plans from part specifications. It suffers from limited dataset availability in industry, reducing model generalization. We propose a semi-supervised learning approac...

📖 Read original article


386. Universal Redundancies in Time Series Foundation Models ​

Author: Anthony Bao, Venkata Hasith Vattikuti, Jeffrey Lai, William Gilpin
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2602.01605v2 Announce Type: replace Abstract: Time Series Foundation Models (TSFMs) leverage extensive pretraining to accurately predict unseen time series during inference, without the need for task-specific fine-tuning. Through large-scale evaluations on standard benchmarks, we find that lea...

📖 Read original article


387. Least but not Last: Fine-tuning Intermediate Principal Components for Better Performance-Forgetting Trade-Offs ​

Author: Alessio Quercia, Arya Bangun, Ira Assent, Hanno Scharr
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2602.03493v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) methods have emerged as crucial techniques for adapting large pre-trained models to downstream tasks under computational and memory constraints. However, they face a fundamental challenge in balancing task-specific perfor...

📖 Read original article


388. Constrained Group Relative Policy Optimization ​

Author: Roger Girgis, Rodrigue de Schaetzen, Luke Rowe, Azal'ee Robitaille, Christopher Pal, Liam Paull
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.RO

arXiv:2602.05863v3 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) remains the dominant critic-free approach for fine-tuning LLMs and VLMs, but its compatibility with constrained policy optimization (e.g. for safety-critical domains) has not been carefully examined. In thi...

📖 Read original article


389. Zero-shot Generalizable Graph Anomaly Detection with Mixture of Riemannian Experts ​

Author: Xinyu Zhao, Qingyun Sun, Jiayi Luo, Xingcheng Fu, Jianxin Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.06859v3 Announce Type: replace Abstract: Graph Anomaly Detection (GAD) aims to identify irregular patterns in graph data, and recent works have explored zero-shot generalist GAD to enable generalization to unseen graph datasets. However, existing zero-shot GAD methods largely ignore intri...

📖 Read original article


390. Reverse N-Wise Output-Oriented Testing for AI/ML and Quantum Computing Systems ​

Author: Lamine Rihani
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2602.14275v2 Announce Type: replace Abstract: Artificial intelligence/machine learning (AI/ML) systems and emerging quantum computing software present unprecedented testing challenges characterized by high-dimensional/continuous input spaces, probabilistic/non-deterministic output distribution...

📖 Read original article


391. Efficient Adaptation of ROMs for Unsteady Flows Using Data Assimilation ​

Author: Isma"el Zighed, Andrea N'ovoa, Luca Magri, Taraneh Sayadi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, physics.flu-dyn

arXiv:2602.23188v3 Announce Type: replace Abstract: We propose an efficient retraining strategy for a parameterized Reduced Order Model (ROM) that attains accuracy comparable to full retraining while requiring only a fraction of the computational time and relying solely on sparse observations of the...

📖 Read original article


392. MUSE: A Run-Centric Platform for Multimodal Unified Safety Evaluation of Large Language Models ​

Author: Zhongxi Wang, Yueqian Lin, Jingyang Zhang, Zinuo Cheng, Hai Helen Li, Yiran Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.CV, cs.SD, eess.AS

arXiv:2603.02482v2 Announce Type: replace Abstract: Safety evaluation of multimodal large language models requires tracking not only whether an attack succeeds, but also how the interaction unfolds across turns and input modalities. We present MUSE (Multimodal Unified Safety Evaluation), an open-sou...

📖 Read original article


393. Personalized Group Relative Policy Optimization for Heterogenous Preference Alignment ​

Author: Jialu Wang, Heinrich Peters, Asad A. Butt, Navid Hashemi, Alireza Hashemi, Pouya M. Ghari, Joseph Hoover, James Rae, Morteza Dehghani
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2603.10009v2 Announce Type: replace Abstract: Despite their sophisticated general-purpose capabilities, Large Language Models (LLMs) often fail to align with diverse individual preferences because standard post-training methods, like Reinforcement Learning with Human Feedback (RLHF), optimize ...

📖 Read original article


394. Group Resonance Network: Learnable Prototypes and Multi-Subject Resonance for EEG Emotion Recognition ​

Author: Renwei Meng
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.11119v2 Announce Type: replace Abstract: Electroencephalography (EEG)-based emotion recognition remains challenging in cross-subject settings due to severe inter-subject variability. Existing methods mainly learn subject-invariant features, but often under-exploit stimulus-locked group re...

📖 Read original article


395. Grammar of the Wave: Towards Explainable Multivariate Time Series Event Detection via Neuro-Symbolic VLM Agents ​

Author: Sky Chenwei Wan, Yifei Y. Wang, Tianjun Hou, Xiqing Chang, Aymeric Jan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.MA

arXiv:2603.11479v4 Announce Type: replace Abstract: Time Series Event Detection (TSED) aims to localize semantically meaningful events in time series data, with critical applications in high-stakes domains. Unlike statistical anomalies, events are often defined by natural-language descriptions with ...

📖 Read original article


396. Generalization and Memorization in Rectified Flow ​

Author: Mingxing Rao, Daniel Moyer
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CV

arXiv:2603.13421v3 Announce Type: replace Abstract: Generative models based on the Flow Matching objective, particularly Rectified Flow, have emerged as a dominant paradigm for efficient, high-fidelity image synthesis. However, while existing research heavily prioritizes generation quality and archi...

📖 Read original article


397. Mixture-Greedy for Online Generative Model Selection: Is UCB Necessary in Diversity-Aware Multi-Armed Bandits? ​

Author: Bahar Dibaei Nia, Farzan Farnia
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2603.21716v2 Announce Type: replace Abstract: Efficient selection among multiple generative models is increasingly important in modern generative AI, where sampling from suboptimal models is costly. This problem can be viewed as a multi-armed bandit (MAB) task. Under diversity-aware evaluation...

📖 Read original article


398. Identification of Bivariate Causal Directionality Based on Anticipated Asymmetric Geometries ​

Author: Alex Glushkovsky
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.LO

arXiv:2603.26024v2 Announce Type: replace Abstract: Identification of causal directionality in bivariate numerical data is a fundamental research problem with important practical implications. This paper presents two alternative methods to identify direction of causation by considering conditional d...

📖 Read original article


399. LIBERO-Para: A Diagnostic Benchmark and Metrics for Paraphrase Robustness in VLA Models ​

Author: Chanyoung Kim, Minwoo Kim, Minseok Kang, Hyunwoo Kim, Dahuin Jung
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2603.28301v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models achieve strong performance in robotic manipulation by leveraging pre-trained vision-language backbones. However, in downstream robotic settings, they are typically fine-tuned with limited data, leading to overfit...

📖 Read original article


400. EvoLen: Evolution-Guided Tokenization for DNA Language Model ​

Author: Nan Huang, Xiaoxiao Zhou, Junxia Cui, Mario Tapia-Pacheco, Tiffany Amariuta, Yang Li, Jingbo Shang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, q-bio.GN

arXiv:2604.08698v2 Announce Type: replace Abstract: Tokens serve as the basic units of representation in DNA language models (DNALMs), yet their design remains underexplored. Unlike natural language, DNA lacks inherent token boundaries or predefined compositional rules, making tokenization a fundame...

📖 Read original article


401. Automated Batch Distillation Process Simulation for a Large Hybrid Dataset for Deep Anomaly Detection ​

Author: Jennifer Werner, Justus Arweiler, Indra Jungjohann, Jochen Schmid, Fabian Jirasek, Hans Hasse, Michael Bortz
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.09166v3 Announce Type: replace Abstract: Anomaly detection (AD) in chemical processes based on deep learning offers significant opportunities but requires large, diverse, and well-annotated training datasets that are rarely available from industrial operations. In a recent work, we introd...

📖 Read original article


402. MOONSHOT : A Framework for Multi-Objective Pruning of Vision and Large Language Models ​

Author: Gabriel Afriat, Xiang Meng, Shibal Ibrahim, Hussein Hazimeh, Rahul Mazumder
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.13287v2 Announce Type: replace Abstract: Weight pruning is a common technique for compressing large neural networks. We focus on the challenging post-training one-shot setting, where a pre-trained model is compressed without any retraining. Existing one-shot pruning methods typically opti...

📖 Read original article


403. REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations ​

Author: Sajjad Ghiasvand, Mark Beliaev, Mahnoosh Alizadeh, Ramtin Pedarsani
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.17289v3 Announce Type: replace Abstract: Supervised fine-tuning of large language models relies on human-annotated data, yet annotation pipelines routinely involve multiple crowdworkers of heterogeneous expertise. Standard practice aggregates labels via majority vote or simple averaging, ...

📖 Read original article


404. Perturbation Sensitivity of Maximum-Likelihood Pairwise Ranking in Computational Decision Systems ​

Author: Junyi Yao, Zihao Zheng, Jiayu Long
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GT

arXiv:2604.17805v3 Announce Type: replace Abstract: Maximum-likelihood pairwise ranking is a com- mon computational mechanism for prioritization, reputation estimation, and comparison-driven decision support. Despite its broad use, the perturbation sensitivity of this estimator under structured chan...

📖 Read original article


405. FG$^2$-GDN: Enhancing Long-Context Gated Delta Networks with Doubly Fine-Grained Control ​

Author: Pingwei Sun, Yuxuan Hu, Jianchao Tan, Xue Wang, Jiaqi Zhang, Yifan Lu, Yerui Sun, Yuchen Xie, Xunliang Cai
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2604.19021v3 Announce Type: replace Abstract: Linear attention mechanisms have emerged as promising alternatives to softmax attention, offering linear-time complexity during inference. Recent advances such as Gated DeltaNet (GDN) and Kimi Delta Attention (KDA) have demonstrated that the delta ...

📖 Read original article


406. Inverting Foundation Models of Brain Function with Simulation-Based Inference ​

Author: Niels Leif Bracher, Xavier Intes, Stefan T. Radev
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, stat.ML

arXiv:2604.23865v3 Announce Type: replace Abstract: Foundation models of brain activity promise a new frontier for in silico neuroscience by emulating neural responses to complex stimuli across tasks and modalities. A natural next step is to ask whether these models can also be used in reverse. Can ...

📖 Read original article


407. AutoREC: A reinforcement learning platform for equivalent circuit model generation ​

Author: Ali Jaberi (Clean Energy Innovation Research Centre, National Research Council Canada, Mississauga, ON, Canada), Yonatan Kurniawan (Department of Materials Science and Engineering, University of Toronto, Toronto, ON, Canada), Robert Black (Clean Energy Innovation Research Centre, National Research Council Canada, Mississauga, ON, Canada), Shayan Mousavi M. (Clean Energy Innovation Research Centre, National Research Council Canada, Mississauga, ON, Canada), Kabir Verma (Cheriton School of Computer Science, University of Waterloo, Waterloo, ON, Canada), Zoya Sadighi (Clean Energy Innovation Research Centre, National Research Council Canada, Mississauga, ON, Canada), Santiago Miret (Lila Sciences, San Francisco, CA, USA), Jason Hattrick-Simpers (Department of Materials Science and Engineering, University of Toronto, Toronto, ON, Canada)
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.mtrl-sci

arXiv:2604.27266v2 Announce Type: replace Abstract: This paper introduces AutoREC, an open-source Python platform for developing, training, and evaluating reinforcement learning (RL) agents that automatically generate equivalent circuit models (ECMs) from electrochemical impedance spectroscopy (EIS)...

📖 Read original article


408. AirFM-DDA: Air-Interface Foundation Model in the Delay-Doppler-Angle Domain for AI-Native 6G ​

Author: Kejia Bian, Meixia Tao, Jianhua Mo, Zhiyong Chen, Leyan Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.IT, eess.SP, math.IT

arXiv:2605.00020v2 Announce Type: replace Abstract: The success of large foundation models is catalyzing a new paradigm for AI-native 6G network design: wireless foundation models for physical-layer design. However, existing models often operate on channel state information (CSI) in the spatial-temp...

📖 Read original article


409. Concepts Whisper: Spectral Anti-Concentration and the Dual Geometry of Transformer Representations ​

Author: Pratyush Acharya, Nuraj Rimal, Habish Dhakal
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.01609v2 Announce Type: replace Abstract: We find that transformer concept representations systematically anti-concentrate in the spectral tail of the unembedding covariance, encoding word-level concepts in low-variance directions across a 17-model core suite and an expanded set of 22 sema...

📖 Read original article


410. Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level ​

Author: Nan Jia, Haojin Yang, Xing Ma, Jiesong Lian, Shuailiang Zhang, Weipeng Zhang, Ke Zeng, Xunliang Cai, Zequn Sun
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.06387v4 Announce Type: replace Abstract: On-policy distillation (OPD) trains a student on its own trajectories with token-level teacher feedback and often outperforms off-policy distillation and standard reinforcement learning. However, we find that its standard advantage weighted policy ...

📖 Read original article


411. PairAlign: A Framework for Autoregressive Tokenization via Self-Alignment with Applications to Audio Tokenization ​

Author: Adhiraj Banerjee, Vipul Arora
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.SD, eess.AS

arXiv:2605.06582v4 Announce Type: replace Abstract: Modern learning systems represent perceptual signals with continuous vectors, but comparison, retrieval, memory, alignment, and reasoning are often symbolic. In language, tokens provide this interface; for speech and audio, it must be learned. Exis...

📖 Read original article


412. OrScale: Orthogonalised Optimization with Layer-Wise Trust-Ratio Scaling ​

Author: Yuxuan Lou, Yang You
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2605.07815v2 Announce Type: replace Abstract: Muon fixes the \emph{direction} of every matrix-valued update at the polar factor of its momentum, while each layer's step \emph{magnitude} is addressed only by a static shape correction. We derive a dynamic per-layer scalar by adapting the LARS/LA...

📖 Read original article


413. ADMM-Q: An Improved Hessian-based Weight Quantizer for Post-Training Quantization of Large Language Models ​

Author: Ryan Lucas, Mehdi Makni, Xiang Meng, Adam Deng, Rahul Mazumder
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.11222v2 Announce Type: replace Abstract: Quantization is an effective strategy to reduce the storage and computation footprint of large language models (LLMs). Post-training quantization (PTQ) is a leading approach for compressing LLMs. Popular weight quantization procedures, including GP...

📖 Read original article


414. Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning ​

Author: Swapnil Parekh, Naman Goyal
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.11467v2 Announce Type: replace Abstract: Reasoning models post-hoc rationalize answers they have already committed to internally, producing chains of reasoning theater: deliberative-looking steps that contribute nothing to correctness. This wastes inference tokens, pollutes interpretabi...

📖 Read original article


415. Emulating the Forced Response of Climate Models with Generative Machine Learning ​

Author: Graham Clyne, Julia Kaltenborn, Peer Nowack, Claire Monteleoni, Anastase Charantonis
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.16929v2 Announce Type: replace Abstract: Global climate models are essential tools to simulate past and potential future pathways of climate change, as well as associated climate impacts. Shared Socioeconomic Pathways (SSPs) describe a range of future scenarios of global economic and demo...

📖 Read original article


416. World Model Control by Trajectory Reachability Metrics ​

Author: Liangyu Li, Shengzhi Wang, Libin Qiu, Mingliang Xiong, Qingwen Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.RO

arXiv:2605.22164v2 Announce Type: replace Abstract: Latent world models can learn representations that contain information needed for control, while the downstream controller may still rank candidate actions poorly when it relies on terminal latent distance alone. We study this failure in a fixed en...

📖 Read original article


417. Generalist Graph Anomaly Detection via Prototype-Based Distillation ​

Author: Yiming Xu, Zihan Chen, Zhen Peng, Song Wang, Bin Shi, Bo Dong, Chao Shen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.26857v2 Announce Type: replace Abstract: Driven by the pressing demand for graph anomaly detection (GAD) in high-stakes domains, the generalist GAD paradigm, which trains a single detector transferable across new graphs, has recently gained growing attention. However, existing methods oft...

📖 Read original article


418. When the Strongest Teacher Is Not the Best Teacher: Student-Centric Answer Selection ​

Author: Zhengyu Hu, Zheyuan Xiao, Linxin Song, Fengqing Jiang, Yuetai Li, Zhihan Xiong, Yue Liu, Junhao Lin, Yao Su, Lijie Hu, Kaize Ding, Teng Xiao, Radha Poovendran
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2605.26872v5 Announce Type: replace Abstract: LLM training increasingly relies on teacher-generated supervision, from synthetic responses to reasoning traces and tool-use demonstrations. Current practice often chooses the highest-performing teacher to generate student training data, implicitly...

📖 Read original article


419. Locality-Aware Redundancy Pruning for LLM Depth Compression ​

Author: Vincent-Daniel Yun, Youngrae Kim, Woosang Lim, YoungJin Heo, Minkyu Kim, Sunwoo Lee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.27786v3 Announce Type: replace Abstract: Large language models are known to contain representational redundancy across network depth, making depth pruning an effective approach for improving inference efficiency. Existing one-shot pruning methods rely on local layer importance or fixed re...

📖 Read original article


420. SYNAPSE: Neuro-Symbolic Visual Thought-to-Text Decoding via Topological Semantic Denoising ​

Author: Akshaj Murhekar, Abhijit Mishra
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2605.27790v2 Announce Type: replace Abstract: Recent advances in large language models have accelerated open-vocabulary EEG-to-imagined-text decoding, where non-invasive neural activity recorded during visual perception is translated into coherent natural language descriptions of viewed stimul...

📖 Read original article


421. OISD: On-Policy Internal Self-Distillation of Language Models ​

Author: Xinyu Liu, Darryl Cherian Jacob, Yang Zhou, Jindong Wang, Pan He
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2605.29089v2 Announce Type: replace Abstract: Recent reinforcement learning (RL) post-training approaches primarily optimize the final output policy using sparse outcome-level rewards, while largely overlooking predictive signals encoded in intermediate representations. In this paper, we intro...

📖 Read original article


422. LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training ​

Author: Minju Gwak, Minseo Kwak, Dongseok Lee, Guijin Son, Alan Ritter, Jaehyung Kim
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2605.29888v2 Announce Type: replace Abstract: Reinforcement learning (RL) post-training has shown to improve reasoning in large language models (LLMs). However, there has been little exploration on the problem of data contamination in RL post-training, potentially undermining generalization an...

📖 Read original article


423. Riemannian Optimization for Hadamard Products of Low-Rank Matrices ​

Author: Pratik Jawanpuria, Ankish Chandresh, Bamdev Mishra
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, math.OC

arXiv:2606.01216v2 Announce Type: replace Abstract: The elementwise Hadamard product of two low-rank matrices provides a parameter-efficient model for data with multiplicative structure, but its modeling is challenging due to the presence of additional symmetries under coupled row/column scalings be...

📖 Read original article


424. Truthful AI Advisors: A Pre-Specified Benchmark for Large Language Model Honesty Under Preference Misalignment ​

Author: Hamidreza Hasani Balyani, Seyed Pouyan Mousavi Davoudi, Alireza Amiri-Margavi, Amin Gholami Davodi, Arshia Gharagozlou
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.GT

arXiv:2606.01456v2 Announce Type: replace Abstract: Large language models are increasingly deployed as advisors whose objective is not aligned with the user's: recommenders optimize for engagement, sales assistants for purchases. Whether they stay truthful when honesty conflicts with their own payof...

📖 Read original article


425. GRZO: Group-Relative Zeroth-Order Optimization for Large Language Model Fine-Tuning ​

Author: Liyan Tan, Yequan Zhao, Yifan Yang, Ruijie Zhang, Xinling Yu, Zheng Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.02857v2 Announce Type: replace Abstract: Zeroth-order (ZO) optimization is a memory-efficient alternative to backpropagation for fine-tuning large language models, but its deployment is limited by the high variance of gradient estimation. We propose GRZO, a Group-Relative Zeroth-Order opt...

📖 Read original article


426. Mamba-Assisted Non-Markovian Closure for Reduced-Order Modeling ​

Author: Zhi-Feng Wei, Saad Qadeer, Panos Stinis
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA, stat.ML

arXiv:2606.05371v2 Announce Type: replace Abstract: Reduced-order modeling of high-dimensional dynamical systems is often hindered by closure effects arising from unresolved variables, which can introduce non-Markovian dependence into the resolved dynamics. Motivated by the history-dependent memory ...

📖 Read original article


427. GRASP: Geometry-aware Residual Alignment for Scalable Pretraining Data Attribution ​

Author: Yue Min, Ruining Chen, Yujun Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.06892v2 Announce Type: replace Abstract: Scalable data attribution methods typically assign isolated utility scores to individual training examples. This prevalent additive assumption fundamentally fails to capture critical subset dynamics, including data redundancy and complementary cove...

📖 Read original article


428. TriHead-GAN: A Generative Adversarial Network with Triple-Head Discriminator for Carbon Emission Time Series Generation ​

Author: Zesen Wang, Lijuan Lan, Yonggang Li, Chunhua Yang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.07569v2 Announce Type: replace Abstract: Accurate carbon emission monitoring is critical for climate policy and emerging regulatory mechanisms such as the EU Carbon Border Adjustment Mechanism, yet city-level high-frequency monitoring data remain extremely scarce, severely limiting data-h...

📖 Read original article


429. QueryGraph: Reliable Multi-Tool Query Execution Planning via LLM-Based Graph Generation ​

Author: Aishwarya Chakravarthy, Vidhi Kulkarni, Duen Horng Chau
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.08300v2 Announce Type: replace Abstract: Many real-world queries over personal data span multiple applications and require structured planning, as individual tools expose only partial information. While LLMs show strong reasoning and tool use, reliably executing multi-step, cross-tool que...

📖 Read original article


430. Divide-and-Conquer Modeling for the CTF-4-Science Lorenz Benchmark ​

Author: Shundong Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.10084v2 Announce Type: replace Abstract: This submission documents the divide-and-conquer modeling strategy developed for the CTF-4-Science Lorenz Chaotic Systems Challenge at AI-DEEDS 2026. The challenge uses the CTF-4-Science Lorenz benchmark to evaluate chaotic-system prediction across...

📖 Read original article


431. When Design Rules Break: Benchmark Composition Determines Whether Label Informativeness Predicts GNN Aggregator Choice ​

Author: Neha Sharma, Ritesh Sharma
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.SI

arXiv:2606.10249v2 Announce Type: replace Abstract: We examine whether graph neural network (GNN) design rules generalize across benchmark families by studying aggregator selection (sum, mean, max) on 24 node-classification datasets spanning citation, heterophilic, LINKX Facebook-100, co-purchase, a...

📖 Read original article


432. Bergson: An Open Source Library for Data Attribution ​

Author: Lucia Quirke, Louis Jaburi, David Johnston, William Z. Li, Gon\c{c}alo Paulo, Guillaume Martres, Girish Gupta, Stella Biderman, Nora Belrose
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.11660v2 Announce Type: replace Abstract: Data attribution is a promising field in interpretability that aims to explain model behavior through the influence of its training data, with applications including debugging undesirable model behavior and training dataset curation. However, signi...

📖 Read original article


433. A Zero-shot Generalized Graph Anomaly Detection Framework via Node Reconstruction ​

Author: Phan Nguyen, Dat Cao, Hien Chu, Khue Hoang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.12673v2 Announce Type: replace Abstract: Cross-domain graph anomaly detection (GAD) aims to identify abnormal nodes in unseen target graphs, showing strong potential in real-world applications with heterogeneous graph data. However, existing methods often depend on dataset-specific featur...

📖 Read original article


434. CARE: Context-Aware Ranking Evolution with Executable Scoring Programs for Budgeted Reaction Optimization ​

Author: Guanyu Liu, Weiyi Kong, Chao Tang, Zeyu Wang, Boer Zhang, Baiqing Li, Peiyu Zhang, Tianyu Shi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2606.14581v5 Announce Type: replace Abstract: High-throughput experimentation can evaluate many reaction conditions, yet combinatorial condition spaces still exceed the available experiment budget. This makes experiment selection a sequential decision problem: each new condition must be chosen...

📖 Read original article


435. Learning Generated Controls under Fractured Geometry: Projective Residualization and Variation-Allocation Frontiers ​

Author: Rui Wu, Zongyuan Chen, Hong Xie, Defu Lian, Enhong Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.14636v3 Announce Type: replace Abstract: Many two-stage estimators assess the first-stage learner by prediction error, even when the next stage uses its residual. In control-function instrumental variables, that residual must preserve the latent control direction without removing the trea...

📖 Read original article


436. WiSP: A Working-Set View of Mixture-of-Experts Serving on Extremely Low-Resource Hardware ​

Author: Jiamu Zhang, Liang Wu, Mayank Darbari, Liangjie Hong
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.21868v2 Announce Type: replace Abstract: Modern local and agentic workloads often need large-model capacity at low concurrency, but run on GPUs that cannot keep a frontier-scale model resident. Mixture-of-Experts (MoE) models are a natural fit because they activate only a small subset of ...

📖 Read original article


437. NeuReasoner: Theory-grounded Mapping of Reasoning Elicitation Boundaries ​

Author: Aydin Javadov, Shyngys Aitkazinov, Tobias Hoesli, Florian von Wangenheim, Bjoern Schuller, Joseph Ollier
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2606.29971v2 Announce Type: replace Abstract: A growing body of work suggests that the reasoning capabilities of large language models are largely latent in their base form, with post-training primarily amplifying rather than introducing them. However, this evidence comes mainly from mathemati...

📖 Read original article


438. Bayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty Estimation ​

Author: Jijie Zhang, Zhe Ren, Quan Zhang, Dandan Guo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.02182v2 Announce Type: replace Abstract: Large language models (LLMs) exhibit remarkable reasoning capabilities, but their task-specific fine-tuning is notoriously plagued by overconfidence, severely hindering trustworthy deployment. We propose Data-Adaptive Lower-Rank Adaptation (DALorRA...

📖 Read original article


439. A Few Teacher Steps Go a Long Way: Cost-Efficient On-Policy Data Augmentation for Agent Post-Training ​

Author: Junze Ye, Jiayi Cheng, Miao Lu, Michal Mankowski, Jose Blanchet, Mohsen Bayati
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.04574v2 Announce Type: replace Abstract: For LLM agents, supervised fine-tuning is not only about teacher labels' quality, but also about which interaction contexts those labels condition on. Pure behavioral cloning uses full teacher demonstrations, creating a mismatch between teacher-ind...

📖 Read original article


440. A Gold-Standard Study of What Makes a Lightweight Game-Playing Agent Strong ​

Author: Nima Kelidari, Mohammadsaeed Haghi, Mahdi Salmani
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.GT

arXiv:2607.06854v2 Announce Type: replace Abstract: Reinforcement learning agents for imperfect-information card games are only as strong as the opponents they train against, and they are hard to grade, since they beat a random opponent over 99 percent of the time and only tie copies of themselves. ...

📖 Read original article


441. Activation Steering Transfer to Agents: One Gain Ratio Does Not Identify Potency and Efficacy ​

Author: Lucas Pinto
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2607.09156v3 Announce Type: replace Abstract: Additive activation steering is calibrated in single-turn chat, then deployed inside agent scaffolds. The quantity usually reported for that move is a gain: a ratio of steered effects, T = Delta_agent / Delta_chat. We sweep eight family x arm dose-...

📖 Read original article


442. PRISM Edit: One Vector for All Temporal Answers ​

Author: Chen Huang, Qi Zheng, Ruiqin Zheng, Long Zeng, Yuantong Xu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11327v3 Announce Type: replace Abstract: Model editing keeps large language models (LLMs) up to date without retraining, but temporal facts expose a limitation of the prevailing locate-and-edit paradigm: an update is not always a replacement. When a fact changes, the new answer should bec...

📖 Read original article


443. Decoupled Structure-Feature Alignment via Alternating Optimization for Graph Learning ​

Author: Chengcheng Yan, Feifei Zhao, Dai Zhu, Wei Liu, Qingsong Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.11577v2 Announce Type: replace Abstract: Conventional Graph Neural Networks (GNNs) couple feature transformation and neighborhood aggregation, which often renders them vulnerable to topological noise and heterophilous connections. To decouple this dependency, we present a constrained two-...

📖 Read original article


444. Energy-Based Physics-Informed Form Finding for Clustered Tensegrity Structures ​

Author: Jing Qin, Muhao Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.NA, math.NA

arXiv:2607.12888v3 Announce Type: replace Abstract: Tensegrity form-finding and physical property prediction are fundamental problems in structural mechanics, which aim to determine equilibrium configurations and internal force distributions. These problems are challenging due to strong nonlinearity...

📖 Read original article


445. Seq2Synth: Benchmarking Temporal Fidelity in Synthetic Sequential Tabular Data ​

Author: Kiwan Kwon, Kangmin Kim, Hojin Lee, Yeseong Jung, Hyeongwoo Kong, Vamsi K. Potluru, Saerom Park, Yongjae Lee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.ML

arXiv:2607.15606v3 Announce Type: replace Abstract: Synthetic sequential tabular data are increasingly used for privacy-preserving data sharing and research, yet conventional tabular metrics often overlook temporal structure. Existing single-table and relational evaluation protocols largely collapse...

📖 Read original article


446. Spaghetti Architect: A Contamination-Resistant, By-Construction-Labelled, Multi-Language Code Dataset Generator ​

Author: Yuxiang Ji
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.SE

arXiv:2607.18642v2 Announce Type: replace Abstract: Mined code corpora are abundant but uncontrolled: a snippet's semantics, surface "messiness," and difficulty are whatever the wild contained; there is no known-optimal reference to grade against; and any public sample may already sit in a model's t...

📖 Read original article


447. Off-Context GRPO: Learning to Reason on Hard Problems using Privileged Information ​

Author: Priyank Agrawal, Ankur Samanta, Shervin Ghasemlou, Boris Vidolov, Jalaj Bhandari, Kavosh Asadi, Daniel Jiang, Aditya Modi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.19313v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) improves reasoning in large language models. Yet, typical RLVR approaches fail on difficult problems: when a model cannot generate any correct solutions, it receives \textit{zero} learning signa...

📖 Read original article


448. Are Single-Token Sparse Autoencoder Features Causally Necessary? Layer-Depth and SAE-Family Effects ​

Author: Seonglae Cho, Zekun Wu, Kleyton Da Costa, Rishi Kalra, Ilham Wicaksono, Adriano Koshiyama
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.20596v2 Announce Type: replace Abstract: Sparse autoencoder (SAE) features are used to interpret and steer large language models, yet nobody has tested whether a feature's causal role is stable across SAE families. Single-token features fire on one vocabulary item, so ground truth permits...

📖 Read original article


449. Error Certificates for KV-Cache Eviction via Randomized Design ​

Author: Peng Xie Amr Alanwar
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2607.21475v3 Announce Type: replace Abstract: Deterministic KV-cache eviction keeps the top-$k$ tokens under an importance score and deletes the rest, and after the deletion the serving system cannot know what the eviction cost it on the current query. We replace the deterministic tail with Po...

📖 Read original article


450. Learning What Matters: Supervising Global Context Pruning with Causal Evidence Sets ​

Author: James E. Allchin
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL

arXiv:2607.21692v4 Announce Type: replace Abstract: Pruning a long context means committing to the blocks a model will keep, and the usual selector is distilled from a dense teacher's attention. That assumes attention shows which context the answer depends on. We test the assumption on retrieval tas...

📖 Read original article


451. Synthetic Speech, Real Signal: Paralinguistic Preservation and Cross-Lingual Augmentation via Voice Cloning ​

Author: Roseline Polle, Owen Parsons, George Fairs, Luis Miguel San Martin Fernandez, Cole Looney, Xiaoliang Wu, Alexandra Livia Georgescu, Stefano Goria
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.SD

arXiv:2607.22304v2 Announce Type: replace Abstract: Synthetic data augmentation in speech is common practice for linguistic tasks like ASR, but has seen far less work for paralinguistic ones, especially clinical tasks where labelled data is expensive and some patient groups are underrepresented. Voi...

📖 Read original article


452. A2TTA: Anchored-and-Agile Test-Time Adaptation for Evolving Traffic Sensor Networks ​

Author: Du Yin, Xiachong Lin, Yue Tan, Jinliang Deng, Estrid He, Hao Xue, Flora D. Salim
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2607.25875v4 Announce Type: replace Abstract: Traffic forecasting is important for efficient traffic management and route planning in smart cities. Existing traffic forecasting studies typically assume fixed sensor graphs, overlooking the continuous evolution of real-world traffic networks, e....

📖 Read original article


453. SERUM: State Extraction and Refinement for User Modeling ​

Author: Andy J. Phu, Karin de Langis, James Mooney, Khanh Chi Le, Dongyeop Kang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV

arXiv:2607.29181v2 Announce Type: replace Abstract: Agentic assistants capable of proactive, personalized interactions require structured models of user intent and workflow. However, building these models from raw, unstructured screen activity remains an open challenge. We present SERUM, a multi-pas...

📖 Read original article


454. An Identifiability Theory of Masked Prediction: Mode Blindness and Mask Schedules ​

Author: Yichao Cai, Javen Qinfeng Shi
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.IT, math.IT, stat.ML

arXiv:2608.01383v3 Announce Type: replace Abstract: Masked prediction learns by inferring missing variables from visible context. This raises a fundamental question: when does near-optimal conditional prediction determine the joint data law? We study this via an $\varepsilon$-identifiability modulus...

📖 Read original article


455. Conformalized Large Language Models under Configuration Shift ​

Author: Yuqicheng Zhu, Jialin Yu, Lin Li, Gengyuan Zhang, Zhen Yang, Steffen Staab, Puneet Dokania, Philip Torr, Jie Tang, Evgeny Kharlamov
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.01460v2 Announce Type: replace Abstract: Conformal prediction (CP) is a distribution-free framework for uncertainty quantification that has recently been adapted to large language models (LLMs), providing prediction sets with finite-sample coverage guarantees under exchangeability. Yet fo...

📖 Read original article


456. Simulation-free and finite-time diffusion model ​

Author: Kentaro Kaba, Masayuki Ohzeki, Yuki Sughiyama
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.03117v2 Announce Type: replace Abstract: The performance of generative diffusion models is determined by the choice of the reference diffusion process connecting the empirical and prior distributions. Conventional approaches typically trade off simulation-free training against finite-time...

📖 Read original article


457. Approximate Speculative Decoding ​

Author: Yuannuo Feng, Zegang Peng, Yuxin Xie, Yubing Ye, Yizhe Chen, Wenshuai Yao, Wenyong Zhou, Wang Kang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.03447v3 Announce Type: replace Abstract: Speculative decoding accelerates autoregressive generation by verifying a draft block with a target model in parallel. Under standard greedy verification, decoding stops at the first draft token that differs from the target argmax, discarding the r...

📖 Read original article


458. CrystalGRPO: Target-Aligned and Coverage-Preserving Reinforcement Learning for Flow-Based Crystal Structure Prediction ​

Author: Kaixiang Su, Hongfei Xue, Qiang Zhu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.06582v2 Announce Type: replace Abstract: Flow-based generative models can efficiently produce candidate structures for crystal structure prediction (CSP), but their pretrained objectives do not directly optimize downstream target recovery. Reinforcement-learning post-training offers a fle...

📖 Read original article


459. From Uncertainty to Failure Attribution: Self-Diagnosing Models for Failure Attribution under Distribution Shift ​

Author: Yiyao Yang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.07953v2 Announce Type: replace Abstract: Distribution shift poses a significant challenge to the robustness of machine learning models, but the current solutions only aim to detect out-of-distribution (OOD) samples and predict uncertainty levels. We introduce a problem setting for failure...

📖 Read original article


460. Support Selection Beyond Smooth DAG Exactness: Completion Geometry,Score Margins, and Selective Certificates ​

Author: Rui Wu, Zongyuan Chen, Hong Xie, Defu Lian, Enhong Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08103v2 Announce Type: replace Abstract: Smooth acyclicity constraints answer whether a weighted support is a DAG, whereas structure learning asks which support change should be made. Existing analyses establish degeneracy for particular constraint formulas but do not isolate what follows...

📖 Read original article


461. Correlation flow governs learning at criticality ​

Author: Andrea Combette, Nelly Pustelnik, Antoine Venaille
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cond-mat.dis-nn

arXiv:2608.08350v2 Announce Type: replace Abstract: The initialization of deep neural networks determines whether information and gradients can propagate across depth, yet a unified theory connecting these properties to learning dynamics remains elusive. Combining mean-field theory and random matrix...

📖 Read original article


462. Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing ​

Author: Srinivasan Manoharan, Junhua Zhao, Fangbo Tu, Haifeng Wu, Jian Wan, Maliah Rajan M, Ashwin Hegde, Mithun Sasidharan, Kalyan Chakravarthi Podamekala
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.08528v2 Announce Type: replace Abstract: Enterprise AI coding assistants incur substantial inference spend, and naive token-cost minimization often fails to reduce end-to-end cost once retries, escalations, and developer wait time are included. We present Task-to-Model Optimization (T2MO)...

📖 Read original article


463. When Do Task Vectors Interfere? Mapping the Validity Boundaries of Weight-Space Composition ​

Author: Chencheng Zhu, Xiaoyang Li, Taotao Cai
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.09490v2 Announce Type: replace Abstract: Task arithmetic composes skills by adding weight displacements, and merged models are then judged on benchmark suites. We measure when that composition is functionally additive, and find that the answer depends as much on how the model is prompted ...

📖 Read original article


464. Terminal Symmetry as a Carrier of Asymmetric Process Knowledge: Statewise Refinement for Anytime Verified Construction ​

Author: Yi Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.11318v2 Announce Type: replace Abstract: Many sequential construction tasks have exact terminal symmetries even though execution is directed and depends on history. Process evidence supplies order; terminal correspondence transports it between equivalent outcomes; the realized state updat...

📖 Read original article


465. Scaling Automatic Research Agents via World Models ​

Author: Xiyuan Yang, Sheikh Sarwar, Jingru Cheng, Zhan Shi, Duanshun Li, Huiyuan Chen, Haiyang Zhang, Xing Fan, Chenlei Guo, Jingrui He, Zhenyu Liao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12564v2 Announce Type: replace Abstract: Automating empirical research is a long-standing direction of AI. Recent automatic research (AutoResearch) agents bring this goal within reach, as modern LLMs show the capability to independently implement solutions and learn from the execution out...

📖 Read original article


466. Momentum as Residual-Driven Multiplier Correction for Deep Learning Optimization ​

Author: Zhixin Ren, Yao Lyu, Congrong Li, Liping Zhang, Shengbo Eben Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.12925v2 Announce Type: replace Abstract: Momentum-based optimizers are widely used in modern deep learning, yet the relations among momentum recursion, update geometry, and acceleration remain only partially understood. We develop an $\textbf{A}$DMM-$\textbf{I}$nspired $\textbf{M}$omentum...

📖 Read original article


467. Reduced Matrix Multiplication: Input-Adaptive Matrix-Product Reduction for LLM Inference ​

Author: Zixuan Lan, Yanhong Li, Jiawei Zhou
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.13426v2 Announce Type: replace Abstract: Transformer-based language models achieve strong performance but incur substantial inference cost due to repeated high-dimensional matrix multiplications. We propose Reduced Matrix Multiplication (RMM), a training-free, input-adaptive inference met...

📖 Read original article


Author: Zhongwei Yu, Yan Song, Xue Yan, Anjie Liu, Xingyu Lu, Yihang Chen, Huichi Zhou, Siyuan Guo, Luoyang Sun, Sihan Chen, Xiangning Yu, Jun Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.15669v2 Announce Type: replace Abstract: Scientific discovery often involves optimising expensive-to-evaluate objectives over vast, structured, and open-ended hypothesis spaces, such as molecules, protein sequences, and computer programs. Generative models such as large language models (L...

📖 Read original article


469. Machine Learning and ARIMA Model Averaging for Adaptive Public Health Forecasting: Comparative Evaluation and an Ontario COVID-19 Case Study ​

Author: Yushu Zou, Ye Li, Johra Moosa, Martin Grunnill, Samir N. Patel, Venkata R. Duvvuri
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, stat.AP

arXiv:2608.20406v2 Announce Type: replace Abstract: Public health forecasts must respond to abrupt changes in surveillance data without over-extrapolating noise, reporting artifacts, or temporary trends. We evaluated autoregressive integrated moving average (ARIMA), random forest, and extreme gradie...

📖 Read original article


470. In-Cell Learning: Language Models That Update Their Own Weights in Sequence Without Changing the File They Ship ​

Author: Zifeng Liu, Yaxin Lu, Xuanhan Wu, Zhiyong Du, Yiming Mao, Zhenhe Wang, Wenqi Shi, Zhengkun Jing, Linwei Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.20873v3 Announce Type: replace Abstract: A 4-bit quantized weight specifies a rounding cell rather than a single full-precision value. We introduce in-cell learning, a paradigm for writing new knowledge only within these cells, so that re-quantizing the served weights reproduces the relea...

📖 Read original article


471. Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data ​

Author: Renfei Zhang, Niloofar Mireshghallah
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.21727v3 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) is deployed to make models better at reasoning tasks, but its side effect on what models will divulge is under studied. Here we show that RLVR on facts increases extraction of personally identif...

📖 Read original article


472. Who Should Teach? Confidence-Aware Dual-Teacher Learning for Few-Shot Node Classification on Text-Attributed Graphs ​

Author: Hojin Kim, Sujin Yoon, Sungsu Lim, Dongwon Lee, David Yoon Suk Kang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.SI

arXiv:2608.22127v2 Announce Type: replace Abstract: Text-Attributed Graphs (TAGs) integrate graph structures and node-associated textual attributes, and recent studies have increasingly leveraged Large Language Models (LLMs) to improve TAG learning in few-shot settings. However, existing approaches ...

📖 Read original article


473. Mol-JEPA: A multimodal Joint Embedding Predictive Architecture for Molecules ​

Author: Florian Rottach, Sebastian Schieferdecker, William Rudman, Randall Balestriero, Carsten Eickhoff
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.22642v3 Announce Type: replace Abstract: Despite recent advances in molecular foundation models, several limitations remain, such as chemically invalid augmentations, modality collapse, and incomplete representation of biochemical environments. To address these challenges, we present \tex...

📖 Read original article


474. CatchBench: When Can an Agent Failure Be Caught? ​

Author: Yue Zhao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.MA, cs.PF

arXiv:2608.22808v2 Announce Type: replace Abstract: When can an agent failure be caught? An audit is usually limited by the record rather than by the method. CatchBench therefore puts one auditor's question to three information states: the declared configuration before a run (PRE), a growing prefix ...

📖 Read original article


475. Every Layer Counts: An Exponential $L_2$ Depth Hierarchy for ReLU Networks ​

Author: Itay Safran
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.23877v2 Announce Type: replace Abstract: We prove a depth hierarchy for ReLU neural networks in which every additional ReLU layer can save exponentially many neurons. For all $k\geq2$, we construct a globally $[0,1]$-valued, $1$-Lipschitz function realized by a depth-$(k+1)$ network of wi...

📖 Read original article


476. Enhancing Bayesian Optimization and Active Learning Through Kernel Diversity ​

Author: Heng Zhang, Haotian Xiang, Konstantinos D. Polyzos, Tara Javidi, Qin Lu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.24721v2 Announce Type: replace Abstract: Hyperparameter selection remains a key challenge in Bayesian optimization (BO) and Bayesian active learning (AL), as model misspecification can lead to suboptimal performance, while more accurate fully Bayesian treatments typically rely on computat...

📖 Read original article


477. FAMPWQ: Fisher Information-based Adaptive Mixed Precision Weight Quantization for Effective LLM Inference ​

Author: Gongwei Lee, Ji Liu, Juncheng Jia, Ji Wu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.DC

arXiv:2608.24945v2 Announce Type: replace Abstract: Recent years have witnessed remarkable achievements of Large Language Models (LLMs) in multiple domains, while the excessive resource requirements of LLMs hinder the deployment on resource-constrained devices. Although model quantization stands out...

📖 Read original article


478. Demystifying Reinforcement Learning Post-Training of Language Models ​

Author: Donovan Clay, Saket Gollapudi, Sankar Harilal, Min Jang, Jacob Morrison, Sewoong Oh, Natasha Jaques
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.24949v2 Announce Type: replace Abstract: Reinforcement learning (RL) post-training has emerged as a powerful framework for enhancing the capabilities of large language models (LLMs), enabling impressive reasoning, math, and coding capabilities. Yet for many researchers and practitioners, ...

📖 Read original article


479. DeMMO: Longitudinal and Cross-Disease Modelling of Digital Mobility Outcomes via Multi-Task Learning ​

Author: Menghui Zhou, Zhipeng Yuan, Vitaveska Lanfranchi, Po Yang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25073v2 Announce Type: replace Abstract: Digital mobility outcomes (DMOs) derived from wearable sensors characterise mobility in daily life and offer a promising means of monitoring disease progression. However, existing DMO studies have typically focused on either a single disease or a s...

📖 Read original article


480. MoPLEx: Estimating Plackett-Luce Mixture Models for Multi-Objective Alignment ​

Author: Dongyue Li, Ziniu Zhang, Lu Wang, Hongyang R. Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CL

arXiv:2608.25200v2 Announce Type: replace Abstract: We study learning a mixture of $k$ Plackett-Luce models from multi-way ranking responses from annotators that may represent heterogeneous underlying preferences. This problem has many applications in AI alignment and preference optimization. Prior ...

📖 Read original article


481. Frequency-aware forecasting for short-term typhoon gust prediction ​

Author: Xuefei Wang, Tingyi Liu, Heng Zhang, Lei Xu, Shengjun Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.25604v2 Announce Type: replace Abstract: Accurate gust forecasting under typhoon conditions remains challenging due to the highly non-stationary and multi-scale characteristics of extreme wind fluctuations. Existing deep learning models often struggle to simultaneously capture long-term t...

📖 Read original article


482. It's a matter of timescale: non-linear utility in successor features and multi-objective planning and learning ​

Author: Liam P. H. Mertens, Lucas N. Alegre, Florent Delgrange, Diederik M. Roijers, Ann Now'e, Peter Vamplew
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.25723v2 Announce Type: replace Abstract: Time is of the essence when dealing with multiple reward signals and non-linear utility. In this paper we argue that the current main approaches in multi-objective RL (SER and ESR), and successor features, are insufficient. While each approach deal...

📖 Read original article


483. ICON Decomposition: Auditing Deep Neural Networks with Multivariate Variance-based Concept-level Explanations ​

Author: Roshan Prakash Rane, Marco Simnacher, Manuel Pfeuffer, Marc-Andre Schulz, Nys Tjade Siegel, Maximilian Dreyer, Frederik Pahde, Wojciech Samek, Sonja Greven, Kerstin Ritter
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI, cs.CV, stat.ML

arXiv:2608.26083v2 Announce Type: replace Abstract: Deep neural networks often exploit spurious associations, a failure known as shortcut learning. Auditing for shortcuts requires testing many candidate concepts, such as acquisition settings or demographics. Current concept-based explainability meth...

📖 Read original article


484. A Unified Framework for Fair and Personalized Decentralized Learning under Communication Constraints ​

Author: Krishnendu S. Tharakan, Carlo Fischione
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.26493v2 Announce Type: replace Abstract: Decentralized learning systems aim to collaboratively train models across multiple clients without relying on a central coordinator. While decentralization improves scalability, privacy, and robustness, it also exacerbates three fundamental challen...

📖 Read original article


485. QuantumBoostNet: Hybrid Classical-Quantum Cardiac View Identification ​

Author: Mihai Udrescu-Milosav, Stefan-Alexandru Jura, Mihai Udrescu, Gerhard-Paul Diller
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG

arXiv:2608.27302v2 Announce Type: replace Abstract: Accurate identification of the correct view or angle in cardiac ultrasound (echocardiogram) is critical for cardiologic imaging, precise anatomical interpretation, and reducing clinical errors. Most state-of-the-art classical models perform well on...

📖 Read original article


486. Beyond Parallel Blindness: Information Floors and Model Gaps in Block Drafting ​

Author: Xinwei Qiang, Xiang Fang, Chang Chen, Yue Guan, Yufei Ding
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.CL, cs.IT, math.IT

arXiv:2608.27339v2 Announce Type: replace Abstract: Block drafters propose several tokens in one forward pass, before earlier target tokens are realised. Their rejection mixes two losses: missing within-block path information and imperfect modelling of observable information. Accepted length cannot ...

📖 Read original article


487. Deriving Scaling Laws for OpenEuroLLM Models: Learning Rate, Batch Size and Loss ​

Author: Niccol`o Ajroldi, Diana Alexandra Onutu, Haider Al-Tahan, J"org Franke, Sampo Pyysalo, Jenia Jitsev, Aaron Klein
Published: 9/1/2026, 4:00:00 AM
Categories: cs.LG, cs.AI

arXiv:2608.28308v2 Announce Type: replace Abstract: We study the scaling behavior of learning rate and batch size in pretraining dense large language models on English-prevalent corpora. Beyond scaling jointly optimal learning rates and batch sizes, we investigate their marginal evolution with model...

📖 Read original article


488. Deep graph kernel point processes over networks ​

Author: Zheng Dong, Matthew Repasky, Xiuyuan Cheng, Yao Xie
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2306.11313v5 Announce Type: replace-cross Abstract: Point process models are widely used for continuous-time discrete-event data, where each data point includes time and additional information called "marks," such as locations, nodes, or event types. We present a new point process model for di...

📖 Read original article


489. Simulation-Based Evaluation of Energy-Constrained Quantum-Classical Competition ​

Author: Junyu Liu, Hansheng Jiang, Zuo-Jun Max Shen
Published: 9/1/2026, 4:00:00 AM
Categories: quant-ph, cs.AI, cs.ET, cs.LG, stat.ML

arXiv:2308.08025v2 Announce Type: replace-cross Abstract: This paper develops a simulation-based framework for evaluating the energy implications of quantum and classical computing firms competing in a market with limited energy resources. We model providers as differentiated Cournot competitors who...

📖 Read original article


490. PhyloGFN: Phylogenetic inference with generative flow networks ​

Author: Mingyang Zhou, Zichao Yan, Elliot Layne, Esmeralda S. Whitammer, Dinghuai Zhang, Moksh Jain, Mathieu Blanchette, Yoshua Bengio
Published: 9/1/2026, 4:00:00 AM
Categories: q-bio.PE, cs.LG, stat.ML

arXiv:2310.08774v3 Announce Type: replace-cross Abstract: Phylogenetics is a branch of computational biology that studies the evolutionary relationships among biological entities. Its long history and numerous applications notwithstanding, inference of phylogenetic trees from sequence data remains c...

📖 Read original article


491. Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning ​

Author: R. Srikant
Published: 9/1/2026, 4:00:00 AM
Categories: math.PR, cs.LG, cs.SY, eess.SY, math.OC

arXiv:2401.15719v5 Announce Type: replace-cross Abstract: We prove a non-asymptotic central limit theorem for vector-valued martingale differences using Stein's method, and use Poisson's equation to extend the result to functions of Markov Chains. We then show that these results can be applied to es...

📖 Read original article


492. PQMass: Probabilistic Assessment of the Quality of Generative Models using Probability Mass Estimation ​

Author: Pablo Lemos, Sammy Sharief, Esmeralda S. Whitammer, Salma Salhi, Connor Stone, Laurence Perreault-Levasseur, Yashar Hezaveh
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, stat.ME

arXiv:2402.04355v4 Announce Type: replace-cross Abstract: We propose a likelihood-free method for comparing two distributions given samples from each, with the goal of assessing the quality of generative models. The proposed approach, PQMass, provides a statistically rigorous method for assessing th...

📖 Read original article


493. Model Selection and Parameter Estimation of One-Dimensional Gaussian Mixture Models ​

Author: Xinyu Liu, Hai Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, eess.SP, stat.ME

arXiv:2404.12613v4 Announce Type: replace-cross Abstract: In this paper, we study the problem of learning one-dimensional Gaussian mixture models (GMMs) with a specific focus on estimating both the model order and the mixing distribution from independent and identically distributed (i.i.d.) samples....

📖 Read original article


494. Learning diverse attacks on large language models for robust red-teaming and safety tuning ​

Author: Seanie Lee, Minsu Kim, Lynn Cherif, David Dobre, Juho Lee, Sung Ju Hwang, Kenji Kawaguchi, Gauthier Gidel, Yoshua Bengio, Esmeralda S. Whitammer, Moksh Jain
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.CR, cs.LG

arXiv:2405.18540v3 Announce Type: replace-cross Abstract: Red-teaming, or identifying prompts that elicit harmful responses, is a critical step in ensuring the safe and responsible deployment of large language models (LLMs). Developing effective protection against many modes of attack prompts requir...

📖 Read original article


495. Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval ​

Author: Kyra Wilson, Aylin Caliskan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CY, cs.AI, cs.CL, cs.LG

arXiv:2407.20371v3 Announce Type: replace-cross Abstract: Artificial intelligence (AI) hiring tools have revolutionized resume screening, and large language models (LLMs) have the potential to do the same. However, given the biases which are embedded within LLMs, it is unclear whether they can be us...

📖 Read original article


496. Autoencoders in Function Space ​

Author: Justin Bunker, Mark Girolami, Hefin Lambley, Andrew M. Stuart, T. J. Sullivan
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2408.01362v4 Announce Type: replace-cross Abstract: Autoencoders have found widespread application in both their original deterministic form and in their variational formulation (VAEs). In scientific applications and in image processing it is often of interest to consider data that are viewed ...

📖 Read original article


497. Classification Drives Geographic Bias in Street Scene Segmentation ​

Author: Rahul Nair, Gabriel Tseng, Esther Rolf, Bhanu Tokas, Hannah Kerner
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.CY, cs.LG

arXiv:2412.11061v2 Announce Type: replace-cross Abstract: Previous studies showed that image datasets lacking geographic diversity can lead to biased performance in models trained on them. While earlier work studied general-purpose image datasets (e.g., ImageNet) and simple tasks like image recognit...

📖 Read original article


498. Perforated Backpropagation: A Neuroscience Inspired Extension to Artificial Neural Networks ​

Author: Rorry Brenner, Laurent Itti
Published: 9/1/2026, 4:00:00 AM
Categories: cs.NE, cs.LG, q-bio.NC

arXiv:2501.18018v2 Announce Type: replace-cross Abstract: The neurons of artificial neural networks were originally invented when much less was known about biological neurons than is known today. Our work explores a modification to the core neuron unit to make it more parallel to a biological neuron...

📖 Read original article


499. An Efficient Sparse Fine-Tuning with Low Quantization Error via Neural Network Pruning ​

Author: Cen-Jhih Li, Aditya Bhaskara
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2502.11439v3 Announce Type: replace-cross Abstract: Fine-tuning is an important step in adapting foundation models such as large language models to downstream tasks. To make this step more accessible to users with limited computational budgets, it is crucial to develop fine-tuning methods that...

📖 Read original article


500. Trust Under Siege: Label Spoofing Attacks against Machine Learning for Android Malware Detection ​

Author: Tianwei Lan, Luca Demetrio, Farid Nait-Abdesselam, Yufei Han, Simone Aonzo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2503.11841v2 Announce Type: replace-cross Abstract: Machine Learning (ML) malware detectors rely heavily on crowd-sourced AntiVirus (AV) labels, with platforms like VirusTotal serving as trusted sources of malware annotations. But what if attackers could manipulate these labels to classify ben...

📖 Read original article


501. Mirror Descent Linearized Augmented Lagrangian Methods for Nonconvex Constrained Stochastic Zeroth-Order Optimization ​

Author: Qiankun Shi, Han Yuan, Xiao Wang, Hao Wang
Published: 9/1/2026, 4:00:00 AM
Categories: math.OC, cs.LG

arXiv:2504.09409v3 Announce Type: replace-cross Abstract: In this paper, we study nonconvex constrained stochastic zeroth-order optimization problems with exact constraints and stochastic objective evaluations. To solve this class of problems, we propose a framework of mirror descent linearized augm...

📖 Read original article


502. mRNA Design and Optimization with Deep Knowledge-Infused Approach ​

Author: Zheng Gong, Ziyi Jiang, Weihao Gao, Yuanyuan Wang, Zhining Cai, Deng Zhuo, Lan Ma
Published: 9/1/2026, 4:00:00 AM
Categories: q-bio.QM, cs.AI, cs.LG

arXiv:2505.23862v2 Announce Type: replace-cross Abstract: The mRNA optimization is essential for mRNA vaccines, therapies, and industrial protein production. Based on current explorations, an ideal optimization approach should simultaneously (i) prevent unintended amino-acid changes, (ii) optimize m...

📖 Read original article


503. Optimal Estimation of Watermark Proportions in Hybrid AI-Human Texts ​

Author: Xiang Li, Garrett Wen, Weiqing He, Jiayuan Wu, Qi Long, Weijie J. Su
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.CL, cs.LG, stat.ME

arXiv:2506.22343v2 Announce Type: replace-cross Abstract: Text watermarks in large language models (LLMs) are an increasingly important tool for detecting synthetic text and distinguishing human-written content from LLM-generated text. While most existing studies focus on determining whether entire ...

📖 Read original article


504. Test of partial effects for Frechet regression on Bures-Wasserstein manifolds ​

Author: Haoshu Xu, Hongzhe Li
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2506.23487v2 Announce Type: replace-cross Abstract: We propose a novel test for assessing partial effects in Fr'echet regression with responses lying on the Bures-Wasserstein manifold. Under the null hypothesis, we show that the statistic admits a degenerate V-statistic approximation whose li...

📖 Read original article


505. PERK: Long-Context Reasoning as Test-Time Learning ​

Author: Zeming Chen, Angelika Romanou, Gail Weiss, Antoine Bosselut
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2507.06415v3 Announce Type: replace-cross Abstract: Long-context reasoning requires accurately identifying relevant information in extensive, noisy input contexts. In this work, we propose PERK (Parameter Efficient Reasoning over Knowledge), a scalable approach for learning to encode long cont...

📖 Read original article


506. Language-Guided Tuning: Configuration Optimization for Automated ML Research ​

Author: Yuxing Lu, Yucheng Hu, Nan Sun, Xukai Zhao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA

arXiv:2508.15757v2 Announce Type: replace-cross Abstract: Configuration optimization remains a critical bottleneck in machine learning, requiring coordinated tuning across model architecture, training strategy, feature engineering, and hyperparameters. Traditional approaches treat these dimensions i...

📖 Read original article


507. Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection ​

Author: Harethah Abu Shairah, Hasan Abed Al Kader Hammoud, George Turkiyyah, Bernard Ghanem
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2508.20766v2 Announce Type: replace-cross Abstract: Safety alignment in Large Language Models (LLMs) often involves mediating internal representations to refuse harmful requests. Recent research has demonstrated that these safety mechanisms can be bypassed by ablating or removing specific repr...

📖 Read original article


508. SPADE: A Large Language Model Framework for Soil Moisture Pattern Recognition and Anomaly Detection in Precision Agriculture ​

Author: Yeonju Lee, Rui Qi Chen, Joseph Oboamah, Po Nien Su, Wei-zhen Liang, Yeyin Shi, Lu Gan, Yongsheng Chen, Xin Qiao, Jing Li
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2509.18123v2 Announce Type: replace-cross Abstract: Accurate interpretation of soil moisture patterns is critical for irrigation scheduling and crop management, yet existing approaches for soil moisture time-series analysis either rely on threshold-based rules or data-hungry machine learning o...

📖 Read original article


509. DCC: Data-Centric Compilation of Machine Learning Kernels for Processing-In-Memory Architectures ​

Author: Peiming Yang, Sankeerth Durvasula, Ivan Fernandez, Mohammad Sadrosadati, Onur Mutlu, Gennady Pekhimenko, Christina Giannoula
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AR, cs.DC, cs.LG, cs.PF

arXiv:2511.15503v5 Announce Type: replace-cross Abstract: High-performance Host processors can integrate Processing-In-Memory (PIM) devices, which can accelerate memory-intensive kernels of Machine Learning (ML) models, including Large Language Models (LLMs), by leveraging the large memory bandwidth...

📖 Read original article


510. TPSO: Training-Free Diverse Image Generation via Semantic Prompt Embedding Optimization ​

Author: Debin Meng, Chen Jin, Zheng Gao, Yanran Li, Ioannis Patras, Georgios Tzimiropoulos
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.CL, cs.LG

arXiv:2511.19811v2 Announce Type: replace-cross Abstract: Image diversity remains a fundamental challenge for text-to-image diffusion models. Low-diversity generation often leads to repetitive outputs, increasing sampling redundancy and hindering both creative exploration and downstream applications...

📖 Read original article


511. SpIDER: Spatially Informed Dense Embedding Retrieval for Software Issue Localization ​

Author: Shravan Chaudhari, Rahul Thomas Jacob, Jiajun Cao, Shihab Rashid, Mononito Goswami, Christian Bock
Published: 9/1/2026, 4:00:00 AM
Categories: cs.SE, cs.LG

arXiv:2512.16956v3 Announce Type: replace-cross Abstract: Retrieving code functions, classes or files relevant to a user query, bug report or feature request from large codebases is a fundamental challenge for Large Language Model (LLM)-based coding agents. Agentic approaches typically employ sparse...

📖 Read original article


512. Diffusion Models in Simulation-Based Inference: A Tutorial Review ​

Author: Jonas Arruda, Niels Bracher, Ullrich K"othe, Jan Hasenauer, Stefan T. Radev
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, stat.ME

arXiv:2512.20685v4 Announce Type: replace-cross Abstract: Diffusion models have recently emerged as powerful learners for simulation-based inference (SBI), enabling fast and accurate estimation of latent parameters from simulated and real data. Their score-based formulation offers a flexible way to ...

📖 Read original article


513. Fitted Q-Evaluation without Bellman Completeness via Occupancy Weighting ​

Author: Lars van der Laan, Nathan Kallus
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2512.23805v4 Announce Type: replace-cross Abstract: Fitted (Q)-evaluation (FQE) is a standard regression-based method for off-policy evaluation, but under distribution shift, value-function realizability alone does not ensure convergence, and existing analyses often require Bellman completen...

📖 Read original article


514. Soft Fitted Q-Iteration without Bellman Completeness: Occupancy Reweighting and Temperature Annealing ​

Author: Lars van der Laan, Nathan Kallus
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2512.23927v3 Announce Type: replace-cross Abstract: Fitted (Q)-iteration (FQI) is a standard regression-based method for optimal control in offline reinforcement learning, but its stability under function approximation often relies on Bellman completeness, which requires Bellman images of th...

📖 Read original article


515. X-Coder: Advancing Competitive Programming with Synthetic Tasks, Solutions, and Tests ​

Author: Jie Wu, Haoling Li, Xin Zhang, Jiani Guo, Jane Luo, Xuewei Yang, Steven Liu, Yangyu Huang, Ruihang Chu, Scarlett Li, Yujiu Yang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2601.06953v3 Announce Type: replace-cross Abstract: Competitive programming remains challenging for code LLMs. Despite recent progress, many training pipelines still depend on scarce real-world data, raising concerns about scalability and near-duplicate benchmark contamination. In this paper, ...

📖 Read original article


516. Tracing the Latent Threads: A Mechanistic Study of How LLMs Represent and Operationalize Race and Ethnicity Cues ​

Author: Shiyue Hu, Ruizhe Li, Yanjun Gao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2601.12868v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly operate in high-stakes settings where demographic attributes such as race and ethnicity may be explicitly stated or implicitly suggested through textual cues. However, existing studies primarily docum...

📖 Read original article


517. Social Caption: Evaluating Social Understanding in Multimodal Models ​

Author: Leena Mathur, Bhaavanaa Thumu, Youssouf Kebe, Louis-Philippe Morency
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2601.14569v3 Announce Type: replace-cross Abstract: Social understanding abilities are crucial for multimodal large language models (MLLMs) to interpret human social interactions. We introduce SOCIAL CAPTION, a framework grounded in interaction theory to evaluate social understanding abilities...

📖 Read original article


518. Learning to Optimize by Differentiable Programming ​

Author: Liping Tao, Xindi Tong, Chee Wei Tan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.MS, cs.LG, math.OC

arXiv:2601.16510v5 Announce Type: replace-cross Abstract: Solving massive-scale optimization problems requires scalable first-order methods with low per-iteration cost. This tutorial highlights a shift in optimization: using differentiable programming not only to execute algorithms but to learn how ...

📖 Read original article


519. Unknown Unknowns: Do Hidden Intentions in LLMs Evade Detection? ​

Author: Devansh Srivastav, David Pape, Lea Sch"onherr
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2601.18552v2 Announce Type: replace-cross Abstract: LLMs expand accessibility and provide wide-reaching access to information. Yet these interactions also create opportunities to embed subtle, goal-oriented behaviours that shape what users think and how they behave, a concern reflected in gove...

📖 Read original article


520. Diverse via bounded Agreement: Geometric Regularization for Multimodal Fusion ​

Author: Zixuan Xia, Hao Wang, Pengcheng Weng, Yanyu Qian, Yangxin Xu, William Dan, Fei Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2601.21670v4 Announce Type: replace-cross Abstract: Multimodal fusion is often treated as an optimization-balancing problem, where training signals are adjusted to prevent one modality from dominating the others. However, balanced optimization does not fully determine the geometry of intermedi...

📖 Read original article


521. Beyond Dense States: Sparse Transcoders as Causally Testable Operators for LLM Latent Reasoning ​

Author: Yadong Wang, Haodong Chen, Yu Tian, Chuanxing Geng, Dong Liang, Xiang Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2602.01695v2 Announce Type: replace-cross Abstract: Latent reasoning reduces the token-generation cost of chain-of-thought reasoning by replacing explicit intermediate tokens with continuous latent transitions. However, existing latent reasoning methods usually rely on dense and entangled tran...

📖 Read original article


522. Efficient reduction of stellar contamination and noise in planetary transmission spectra using neural networks ​

Author: David S. Duque-Casta~no, Lauren Flor-Torres, Jorge I. Zuluaga
Published: 9/1/2026, 4:00:00 AM
Categories: astro-ph.EP, astro-ph.IM, cs.LG

arXiv:2602.10330v4 Announce Type: replace-cross Abstract: The characterization of exoplanetary atmospheres has been transformed by the James Webb Space Telescope (JWST), whose infrared sensitivity enables transmission spectroscopy at unprecedented precision. However, stellar heterogeneities (e.g., s...

📖 Read original article


523. Edge-Local and Qubit-Efficient Quantum Graph Learning for the NISQ Era ​

Author: Armin Ahmadkhaniha, Jake Doliskani
Published: 9/1/2026, 4:00:00 AM
Categories: quant-ph, cs.ET, cs.LG

arXiv:2602.16018v4 Announce Type: replace-cross Abstract: Graph neural networks (GNNs) are a powerful framework for learning representations from graph-structured data, but their direct implementation on near-term quantum hardware remains challenging due to circuit depth, multi-qubit interactions, a...

📖 Read original article


524. AI-Generated Measurements for Identification and Inference with Missing Data: A Weak Shadow Variable Approach ​

Author: Hongyu Chen, David Simchi-Levi, Ruoxuan Xiong
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, econ.EM, stat.ME

arXiv:2602.16061v3 Announce Type: replace-cross Abstract: Across business and social science applications, outcomes are often missing in ways that depend on the unobserved outcomes themselves. In service systems, for example, whether a customer submits a rating depends on the rating they would have ...

📖 Read original article


525. DesignAsCode: Bridging Structural Editability and Visual Fidelity in Graphic Design Generation ​

Author: Ziyuan Liu, Shizhao Sun, Danqing Huang, Yingdong Shi, Meisheng Zhang, Ji Li, Jingsong Yu, Jiang Bian
Published: 9/1/2026, 4:00:00 AM
Categories: cs.GR, cs.AI, cs.CV, cs.LG, cs.MM

arXiv:2602.17690v3 Announce Type: replace-cross Abstract: Graphic design generation demands a delicate balance between high visual fidelity and fine-grained structural editability. However, existing approaches typically bifurcate into either non-editable raster image synthesis or abstract layout gen...

📖 Read original article


526. Prediction-Powered Conditional Inference ​

Author: Yang Sui, Jin Zhou, Hua Zhou, Xiaowu Dai
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2603.05575v2 Announce Type: replace-cross Abstract: We study prediction-powered conditional inference in the setting where labeled data are scarce, unlabeled covariates are abundant, and a black-box machine-learning predictor is available. The goal is to perform statistical inference on condit...

📖 Read original article


527. NanoVDR: Distilling a 2B Vision-Language Retriever into a 70M Text-Only Encoder for Visual Document Retrieval ​

Author: Zhuchenyang Liu, Yao Zhang, Yu Xiao
Published: 9/1/2026, 4:00:00 AM
Categories: cs.IR, cs.CV, cs.LG

arXiv:2603.12824v3 Announce Type: replace-cross Abstract: Vision-Language Model (VLM) based retrievers have advanced visual document retrieval (VDR) to impressive quality. They require the same multi-billion parameter encoder for both document indexing and query encoding, incurring high latency and ...

📖 Read original article


528. PA3: Policy-Aware Agent Alignment through Chain-of-Thought ​

Author: Shubhashis Roy Dipta, Daniel Bis, Kun Zhou, Lichao Wang, Benjamin Z. Yao, Chenlei Guo, Ruhi Sarikaya
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2603.14602v3 Announce Type: replace-cross Abstract: Conversational assistants powered by large language models (LLMs) excel at tool-use tasks but struggle with adhering to complex, business-specific rules. While models can reason over business rules provided in context, including all policies ...

📖 Read original article


529. Interpretable Predictability-Based AI Text Detection: A Replication Study ​

Author: Adam Skurla, Dominik Macko, Jakub Simko
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2603.15034v2 Announce Type: replace-cross Abstract: This paper replicates and extends the system used in the AuTexTification shared task for authorship attribution of machine-generated texts. Exact replication was not possible because of differences in data splits, model availability, and impl...

📖 Read original article


530. Model Selection and Parameter Estimation for Multidimensional Gaussian Mixture Models with a Common Covariance Matrix ​

Author: Xinyu Liu, Hai Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG

arXiv:2603.19657v2 Announce Type: replace-cross Abstract: We study model-order selection and component-mean estimation for multidimensional Gaussian mixture models with a known common covariance matrix. Using empirical characteristic-function measurements, we construct Fourier covariance matrices wh...

📖 Read original article


531. Accelerate Vector Diffusion Maps by Landmarks ​

Author: Sing-Yuan Yeh, Yi-An Wu, Hau-Tieng Wu, Mao-Pei Tsui
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.LG, math.DG, physics.data-an

arXiv:2603.21247v2 Announce Type: replace-cross Abstract: We propose a landmark-constrained algorithm, LA-VDM (Landmark Accelerated Vector Diffusion Maps), to accelerate the Vector Diffusion Maps (VDM) framework built upon the Graph Connection Laplacian (GCL), which captures pairwise connection rela...

📖 Read original article


532. PeopleSearchBench: Evaluating AI-Powered People Search Platforms with Criteria-Grounded Verification ​

Author: Tianyu Shi, Wei Wang, Zequn Xie, Shuai Zhang, Boyang Xia, Chenyu Zeng, Qi Zhang, Lynn Ai, Yaqi Yu, Kaiming Zhang, Feiyue Tang, Zhenyu Yu, Lei Ding
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2603.27476v3 Announce Type: replace-cross Abstract: AI-powered people search platforms are increasingly deployed for recruiting, sales prospecting, and professional networking, yet no standardized benchmark exists for their rigorous evaluation. We present PeopleSearchBench, an open-source benc...

📖 Read original article


533. Robust Multi-Agent Reinforcement Learning for Small UAS Separation Assurance under GPS Degradation and Spoofing ​

Author: Alex Zongo, Filippos Fotiadis, Ufuk Topcu, Peng Wei
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG, cs.SY, eess.SY

arXiv:2603.28900v3 Announce Type: replace-cross Abstract: We address robust separation assurance for small Unmanned Aircraft Systems (sUAS) under GPS degradation and spoofing via Multi-Agent Reinforcement Learning (MARL). In cooperative surveillance, each aircraft (or agent) broadcasts its GPS-deriv...

📖 Read original article


534. Operator Learning for Predicting Bulk Wave Parameters of Spectral Wave Models ​

Author: Shukai Cai, Sourav Dutta, Mark Loveland, Eirik Valseth, Peter Rivera-Casillas, Corey Trahan, Clint Dawson
Published: 9/1/2026, 4:00:00 AM
Categories: physics.comp-ph, cs.LG, physics.flu-dyn

arXiv:2604.06433v2 Announce Type: replace-cross Abstract: The impact of wave-induced forcing on the mean water level and nearshore currents is typically modeled through excess momentum fluxes, also known as radiation stresses, and their spatial gradients. Accurate storm surge prediction requires cou...

📖 Read original article


535. SynMulti: Synthetic-to-Real Learning for Multimodal Video Understanding ​

Author: Tanzila Rahman, Renjie Liao, Leonid Sigal
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2604.12335v2 Announce Type: replace-cross Abstract: Training multimodal large language models (MLLMs) for video understanding requires large-scale annotated data spanning diverse tasks such as object counting, question answering, and segmentation. However, collecting and annotating multimodal ...

📖 Read original article


536. VeriX-Anon: A Multi-Layered Framework for Mathematically Verifiable Outsourced Target-Driven Data Anonymization ​

Author: Miit Daga, Swarna Priya Ramu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CR, cs.DB, cs.LG

arXiv:2604.12431v3 Announce Type: replace-cross Abstract: Organisations increasingly outsource privacy-sensitive data transformations to cloud providers, yet no practical mechanism lets the data owner verify that the contracted algorithm was faithfully executed. VeriX-Anon is a multi-layered verific...

📖 Read original article


537. Performance Manipulation: Labor Market Implications in AI-assisted Era ​

Author: Xiaoyun Qiu, Yang Yu, Haifeng Xu
Published: 9/1/2026, 4:00:00 AM
Categories: econ.GN, cs.GT, cs.LG, q-fin.EC

arXiv:2604.22230v2 Announce Type: replace-cross Abstract: Performance manipulation arises when agents exploit easily measurable, routine tasks to inflate observable outcomes without contributing genuine innovation or expert judgment. We formalize this phenomenon in a game-theoretic model in which ag...

📖 Read original article


538. Large language model-enabled automated data extraction for concrete materials informatics ​

Author: Zhanzhao Li, Kengran Yang, Qiyao He, Kai Gong
Published: 9/1/2026, 4:00:00 AM
Categories: cond-mat.mtrl-sci, cs.CL, cs.LG

arXiv:2604.22938v3 Announce Type: replace-cross Abstract: The promise of data-driven materials discovery remains constrained by the scarcity of large, high-quality, and accessible experimental datasets. Here, we introduce a generalizable large language model (LLM)-powered pipeline for automated extr...

📖 Read original article


539. When Chain-of-Thought Fails, the Solution Hides in the Hidden States ​

Author: Houman Mehrafarin, Amit Parekh, Ioannis Konstas
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2604.23351v2 Announce Type: replace-cross Abstract: Whether intermediate reasoning is computationally useful or merely explanatory depends on whether chain-of-thought (CoT) tokens contain task-relevant information. We present a mechanistic causal analysis of CoT on GSM8K using activation patch...

📖 Read original article


540. Learning the Channel Gain from Anywhere to Anywhere via Cross-environment Transformer Estimators ​

Author: Prasenjit Dhara, Daniel Romero
Published: 9/1/2026, 4:00:00 AM
Categories: eess.SP, cs.IT, cs.LG, math.IT

arXiv:2605.08211v2 Announce Type: replace-cross Abstract: Channel-gain maps provide the channel gain between any two locations in a geographical region. They find numerous applications, from resource allocation and interference control to path planning for autonomous vehicles. Channel-gain map estim...

📖 Read original article


541. Robust Multi-Agent LLMs under Byzantine Faults ​

Author: Haejoon Lee, Vincent-Daniel Yun, Dimitra Panagou, Sai Praneeth Karimireddy
Published: 9/1/2026, 4:00:00 AM
Categories: cs.MA, cs.AI, cs.LG

arXiv:2605.09076v3 Announce Type: replace-cross Abstract: Large language model (LLM) agents increasingly collaborate over peer-to-peer networks to improve their reliability. However, these same interactions can also introduce vulnerability to unreliable or Byzantine agents that can propagate incorre...

📖 Read original article


542. Operator-Guided Model Reduction for Generative Sampling in Lattice Field Theory ​

Author: Moxian Qian
Published: 9/1/2026, 4:00:00 AM
Categories: hep-lat, cs.LG

arXiv:2605.11199v2 Announce Type: replace-cross Abstract: Neural generative samplers for lattice field theory can be costly to train and evaluate. When they miss modes or assign them incorrect relative weights, biased observables do not reveal which collective variables are responsible. We project a...

📖 Read original article


543. Making the Discrete Continuous: Synthetic RAW Augmentations for Fine-Grained Evaluation of Person Detection Performance in Low Light ​

Author: Valeria Pais, Malena Mendilaharzu, Daniele Faccio, Luis Oala, Christoph Clausen, Bruno Sanguinetti
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG, physics.optics

arXiv:2605.22455v2 Announce Type: replace-cross Abstract: Real-world deployment of AI vision models is both fueled and limited by the data available for training and testing. Real datasets are sparse and uneven: long-tailed or unbalanced distributions hinder generalization, and the low number of sam...

📖 Read original article


544. SciAtlas: A Computable Atlas of Science for Knowledge-Grounded AI Research ​

Author: Shuofei Qiao, Yunxiang Wei, Busheng Zhang, Mengru Wang, Jiazheng Fan, Huadong Jian, Bin Wu, Shumin Deng, Yida Xue, Zifan Cheng, Xiang Chen, Dan Zhang, Junfeng Fang, Ningyu Zhang, Keyan Ding, Qiang Zhang, Jeff Z. Pan, Emine Yilmaz, Huajun Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.IR, cs.LG

arXiv:2605.22878v2 Announce Type: replace-cross Abstract: Artificial intelligence is rapidly entering the core workflows of scientific research. Yet reliable scientific reasoning requires access to accumulated scientific knowledge with sufficient breadth, depth, and standardization. Current AI scien...

📖 Read original article


545. Physics-Guided Concentration Inference from Resistance Transients in a Mixed-Phase SnO-SnO$_2$ Carbon Monoxide Sensor with p-n Switching ​

Author: Sani Biswas, Preetam Singh, Amit Kumar Gangwar
Published: 9/1/2026, 4:00:00 AM
Categories: physics.chem-ph, cs.LG, physics.app-ph

arXiv:2605.23971v3 Announce Type: replace-cross Abstract: This work presents a physics-guided machine-learning framework for carbon monoxide concentration inference from experimentally measured resistance transients of a mixed-phase SnO-SnO$_2$ material gas sensor exhibiting temperature-dependent p-...

📖 Read original article


546. Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models ​

Author: Weixian Waylon Li, Mengyu Wang, Tiejun Ma
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CE, cs.LG

arXiv:2605.24564v2 Announce Type: replace-cross Abstract: Backtesting large language models (LLMs) on historical financial data is unreliable when their pre-training data include the evaluated events. An LLM trained in 2024 may already encode how stocks moved during 2018-2020. We name this failure p...

📖 Read original article


547. Measuring the Depth of LLM Unlearning via Activation Patching ​

Author: Jaeung Lee, Dohyun Kim, Jaemin Jo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2605.24614v2 Announce Type: replace-cross Abstract: Large language model (LLM) unlearning has emerged as a crucial post-hoc mechanism for privacy protection and AI safety, yet auditing whether target knowledge is truly erased remains challenging. Existing output-level metrics fail to detect wh...

📖 Read original article


548. Correcting test set contamination by spiking the training data ​

Author: Johnny Tian-Zheng Wei, Jerry Li, Ameya Godbole, Robin Jia
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ME, cs.CL, cs.LG

arXiv:2605.24818v3 Announce Type: replace-cross Abstract: The literature on test set contamination largely focuses on detection, but the correction of contaminated test scores is underexplored. Our core proposal is to spike the training data by intentionally contaminating some test examples at known...

📖 Read original article


549. Universal Activation Verbalizer: A Unified Framework for Cross-Model Activation Explanation ​

Author: Haiyan Zhao, Zirui He, Guanchu Wang, Ali Payani, Yingcong Li, Mengnan Du
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2605.25903v2 Announce Type: replace-cross Abstract: Activation verbalization explains hidden representations in natural language, but existing methods are mostly limited to self-explanation, where each model explains only its own activations. We introduce Universal Activation Verbalizer (UAV),...

📖 Read original article


550. Zipping the Thought: When and How Compressed Reasoning Data Works in LLM Post-Training ​

Author: Kohsei Matsutani, Gouki Minegishi, Takeshi Kojima, Yusuke Iwasawa, Yutaka Matsuo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2605.28008v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can now solve complex problems through long chain-of-thought (CoT) reasoning, but the trade-off between performance and token cost remains a central challenge. To address this issue, supervised fine-tuning (SFT) o...

📖 Read original article


551. Extracting Small Translation Specialists from LLMs by Aggressively Pruning Experts ​

Author: Liu O. Martin, Lucas Bandarkar, Nanyun Peng
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2605.28042v2 Announce Type: replace-cross Abstract: Modern large language models (LLMs) achieve state-of-the-art machine translation performance, but they do so as broad generalists largely trained for many tasks and capabilities unrelated to translation. Thus, they are heavily overparameteriz...

📖 Read original article


552. Privacy-Enhanced Zero-Order Federated Learning via xMK-CKKS over Wireless Channels ​

Author: Anthony Ayli, Khalil Harris, Jihad Fahs, Mohamad Assaad
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CR, cs.LG

arXiv:2605.30123v3 Announce Type: replace-cross Abstract: Homomorphic encryption (HE) enables privacy-preserving aggregation in federated learning (FL) by allowing the server to operate on encrypted data without decryption. Existing HE-over-the-air (OTA) methods mainly rely on single-key HE schemes ...

📖 Read original article


553. When Should Models Change Their Minds? Contextual Belief Management in Large Language Models ​

Author: Haoming Xu, Weihong Xu, Zongrui Li, Mengru Wang, Yunzhi Yao, Chiyu Wu, Jin Shang, Yu Gong, Shumin Deng
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2605.30219v2 Announce Type: replace-cross Abstract: Long-horizon interactions require language models to manage accumulating information: when to update their state, when to preserve their state, and what to ignore. We study this challenge as Contextual Belief Management (CBM): maintaining a p...

📖 Read original article


554. Exploring Autonomous Agentic Data Engineering for Model Specialization ​

Author: Yujie Luo, Xiangyuan Ru, Jingsheng Zheng, Jingjing Wang, Yuqi Zhu, Jintian Zhang, Runnan Fang, Kewei Xu, Ye Liu, Zheng Wei, Jiang Bian, Zang Li, Shumin Deng
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.IR, cs.LG

arXiv:2605.30407v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated strong performance on general tasks, while often struggling to adapt to specialized domains without high-quality domain-specific data. Existing LLM-based data curation methods primarily rely on h...

📖 Read original article


555. Self-Correction Can Amplify Hallucinations: Fact-Level Repair with Graph-Based Evidence Routing in Multimodal Generation ​

Author: Kaixiang Zhao, Tianrun Yu, Shawn Huang, Porter Jenkins, Yushun Dong, Amanda Hughes
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.00232v2 Announce Type: replace-cross Abstract: We study fact-level repair for multimodal generation, where a fluent output may contain specific facts that are not supported by the input. Existing inference-time repair methods often generate feedback by jointly conditioning on the input an...

📖 Read original article


556. DASH: Dual-Branch Score Distillation for Guidance-Calibrated Compact Diffusion Models ​

Author: Abdullah Al Shafi, Kazi Saeed Alam, Sk Imran Hossain, Engelbert Mephu Nguifo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2606.00798v2 Announce Type: replace-cross Abstract: Parameter compression of class-conditional diffusion models exposes a structural limitation in output-level distillation: supervising only the guided output leaves the two score branches non-identifiable, so the classifier-free guidance gap i...

📖 Read original article


557. Don't Read Everything: A Curvature-Conditioned Query for Linear Attention ​

Author: Dong Le, Thong Nguyen, Cong-Duy Nguyen, Anh Tuan Luu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2606.01294v2 Announce Type: replace-cross Abstract: Linear attention reduces the quadratic cost of softmax attention by maintaining a recurrent fast-weight state, but it consistently lags on in-context retrieval and long-context tasks. Existing remedies act on the write side of memory through ...

📖 Read original article


558. Evolving Agents in the Dark: Retrospective Harness Optimization via Self-Preference ​

Author: Wenbo Pan, Shujie Liu, Chin-Yew Lin, Jingying Zeng, Xianfeng Tang, Xiangyang Zhou, Yan Lu, Xiaohua Jia
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2606.05922v3 Announce Type: replace-cross Abstract: AI agents rely on a harness of skills, tools, and workflows to solve complex problems. Continually improving this harness is essential for adapting to new tasks. However, existing optimization methods typically require ground-truth validation...

📖 Read original article


559. CrowdMath: A Dataset of Crowdsourced Mathematical Research Discussions ​

Author: Sherin Muckatira, Jesse Geneson, Slava Gerovitch, Pavel Etingof, Mikhail Gronas, Anna Rumshisky
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.06526v2 Announce Type: replace-cross Abstract: Large language models have made substantial progress on mathematical reasoning, but existing benchmarks typically evaluate well-specified problems with final answers, step-by-step solutions, or complete proofs. They do not capture collaborati...

📖 Read original article


560. Twelve quick tips for designing AI-driven HPC workflows ​

Author: Jamie J. Alnasir
Published: 9/1/2026, 4:00:00 AM
Categories: cs.DC, cs.AI, cs.LG, cs.SE

arXiv:2606.07491v2 Announce Type: replace-cross Abstract: High-performance computing (HPC) clusters remain the backbone of large-scale scientific computation, traditionally executing deterministic, linear pipelines optimised for predictable performance. However, the pervasive integration of artifici...

📖 Read original article


561. How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions ​

Author: Donghao Huang, Tomas Drietomsky, Benjamin Barrett, Zhaoxia Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2606.08051v3 Announce Type: replace-cross Abstract: Merchant information extraction turns noisy financial transaction descriptors into structured fields at production scale. Our deployed LoRA-fine-tuned LLaMA~3.1-8B reaches 96.95% F1, but its memory and throughput motivate smaller replacement...

📖 Read original article


562. From AGI to ASI ​

Author: Tim Genewein, Matija Franklin, Alexander Lerchner, Laurent Orseau, Samuel Albanie, Adam Bales, Cole Wyeth, Stephanie Chan, Iason Gabriel, Joel Z. Leibo, Allan Dafoe, Marcus Hutter, Thore Graepel, Shane Legg
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CY, cs.LG

arXiv:2606.12683v2 Announce Type: replace-cross Abstract: Over the last decade, building human-level artificial general intelligence has moved from far-fetched speculation to being a concrete next-decade target for many of the largest AI organisations. Achieving this goal would have profound and far...

📖 Read original article


563. PEAR: Permutation-Equivariant Adaptive Routing Multi-Agent Debate ​

Author: Yang Feng, Ziwei Xu, Xia Hu, Fengxiang He
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, cs.MA, stat.ML

arXiv:2606.20621v2 Announce Type: replace-cross Abstract: Multi-agent debate improves the reliability of large language models (LLMs) through iterative peer critiques. However, fixed topologies often introduce persistent positional biases, amplify unreliable agents, and cause high sensitivity to rol...

📖 Read original article


564. In LLM Reasoning, there is Irrationality on top of Value Misalignment ​

Author: Kejiang Qian, Fengxiang He
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG, stat.ML

arXiv:2606.20624v2 Announce Type: replace-cross Abstract: Significant progress has been made in aligning LLMs with target value functions. We argue that, even when an LLM has been well aligned in (post-)training, it may still fail to maximise the aligned value in reasoning. We mathematically formali...

📖 Read original article


565. FracEvent: Event-Camera Simulation via Fractional-Relaxation Pixel Dynamics ​

Author: Langyi Chen, Chuanzhi Xu, Haoxian Zhou, Pengfei Ye, Ziyu Luo, Haodong Chen, Qiang Qu, Xiaoming Chen, Weidong Cai
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV

arXiv:2606.26636v2 Announce Type: replace-cross Abstract: Event cameras asynchronously report brightness changes with microsecond-level temporal resolution, but real event data remain difficult to collect at scale because specialized sensors, careful synchronization, and task-specific annotations ar...

📖 Read original article


566. MultiHashFormer: Hash-based Generative Language Models ​

Author: Huiyin Xue, Atsuki Yamaguchi, Nikolaos Aletras
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.AI, cs.LG

arXiv:2606.28057v2 Announce Type: replace-cross Abstract: Language models (LMs) represent tokens using embedding matrices that scale linearly with the vocabulary size. To constrain the parameter footprint, prior work proposes hashing many tokens into a single vector within encoder-only models. While...

📖 Read original article


567. J-LAW: Joint Localization and Action-Conditioned World Modeling via Coupled Latent Factor Graphs ​

Author: Guanqun Cao, Liang Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.LG

arXiv:2606.28712v2 Announce Type: replace-cross Abstract: Classical simultaneous localization and mapping (SLAM) estimates metric poses and a geometric map but does not provide an action-conditioned predictive state. Action-conditioned world models learn compact latent dynamics but ignore global met...

📖 Read original article


568. Self-Organized Conformal Prediction: Reducing Regional Coverage Gaps with Unsupervised Group Discovery ​

Author: Louis Berthier, Ahmed Shokry, Maxime Moreaud, Guillaume Ramelet, Aymeric Dieuleveut
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG

arXiv:2606.29403v2 Announce Type: replace-cross Abstract: Conformal prediction guarantees marginal coverage, but a pooled calibration quantile can hide systematic undercoverage across heterogeneous regions of the feature space. We introduce Self-Organized Conformal Prediction (SOCP), a calibration s...

📖 Read original article


569. Cultural Bias Without a Cultural Self:A Disassociation Study of LLM's Persona and Bias ​

Author: Yuan Yuan
Published: 9/1/2026, 4:00:00 AM
Categories: stat.ML, cs.AI, cs.LG, math.DG

arXiv:2607.02368v3 Announce Type: replace-cross Abstract: Language models prompted with cultural personas increasingly stand in for human respondents in cross-cultural research. Their responses separate personas cleanly, and that separation is read as evidence of a cultural point of view. We show th...

📖 Read original article


570. Runtime Safety Filtering for Learned Small UAS Separation Policies under GNSS Degradation ​

Author: Alex Zongo, Peng Wei
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.LG, cs.MA, cs.SY, eess.SY

arXiv:2607.10014v2 Announce Type: replace-cross Abstract: Learning-based separation assurance for small Unmanned Aircraft Systems (sUAS) achieves near-zero collision rates in simulation, but assumes accurate position and velocity information from Global Navigation Satellite Systems (GNSS). This assu...

📖 Read original article


571. Posterior Variance Is a Constraint Map, Not an Error Map: Closed-Form Uncertainty for Radiative Gaussian Splatting in Sparse-View CT ​

Author: Chulin Zhao, Yiran Xu, Shu Liu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG, eess.IV

arXiv:2607.13682v3 Announce Type: replace-cross Abstract: Radiative Gaussian splatting reconstructs sparse-view CT fast and accurately, and recent work attaches per-Gaussian posteriors to yield per-voxel uncertainty maps. We ask what such a map actually measures: posterior variance is a data-constra...

📖 Read original article


572. RegionFM: Interpretable Region-Based Brain MRI Classification Using Foundation Model Embeddings ​

Author: Wei Zhang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.LG

arXiv:2607.16325v2 Announce Type: replace-cross Abstract: Foundation models provide powerful representations for brain MRI analysis, but their predictions remain difficult to interpret in anatomically meaningful terms. Clinical assessment of brain MRI is commonly organized around anatomically define...

📖 Read original article


573. The Value of Depth in Message Passing on Sparse Graphs: A Kesten-Stigum Dichotomy ​

Author: Aseem Raj Baranwal
Published: 9/1/2026, 4:00:00 AM
Categories: math.ST, cs.LG, math.PR, stat.ML, stat.TH

arXiv:2607.16676v2 Announce Type: replace-cross Abstract: How deep does a graph neural network need to be on a sparse graph? We study its purest statistical form: node classification on the sparse contextual stochastic block model (CSBM) with average degree $\Delta=O(1)$, whose local weak limit is a...

📖 Read original article


574. SALT: Salience-Aware Lexical Trie for Long-Context Compression ​

Author: Oteo Mamo, Hyunjin Yi, Joydhriti Choudhury, Shangqian Gao, Weikuan Yu
Published: 9/1/2026, 4:00:00 AM
Categories: cs.PF, cs.AI, cs.LG

arXiv:2607.17486v2 Announce Type: replace-cross Abstract: As large language models (LLMs) process increasingly longer prompts, computation and KV-cache memory costs have emerged as major bottlenecks in inference systems. Existing input-level prompt compression methods address this, but rank each sen...

📖 Read original article


575. WorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Football Forecasting ​

Author: Zhaokai Wang, Tianlin Gui, Jiayuan Rao, Shangzhe Di, Yihong Tang, Dingli Liang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.CL, cs.LG

arXiv:2607.18084v2 Announce Type: replace-cross Abstract: Predicting a football match before kickoff requires more than knowing past results: a model must use changing information and make a clear prediction before the answer is available. We present WorldCupArena, a dynamic benchmark for language m...

📖 Read original article


576. Action from Adjacent Set in Physical Space Outperforms the Best Prediction in World Models ​

Author: Liangyu Li, Qingwen Liu, Mingqing Liu, Wen Fang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.LG

arXiv:2607.23602v2 Announce Type: replace-cross Abstract: Controllers based on sampling and latent world models assign a predicted terminal cost to each candidate action sequence, choose the minimum, execute its first action block, and replan. This rule can fail even when the terminal cost perfectly...

📖 Read original article


577. Learning to Trace Seiberg Dualities ​

Author: Jonathan J. Heckman, Shani Meynet, Alessandro Mininno, Gary Shiu
Published: 9/1/2026, 4:00:00 AM
Categories: hep-th, cs.AI, cs.LG, hep-ph

arXiv:2607.28628v2 Announce Type: replace-cross Abstract: Dualities play an important role in establishing both microscopic and emergent phenomena in a wide range of physical systems. In practice, though, it can often be computationally challenging to establish when two systems are dual, even when a...

📖 Read original article


578. Fast Trainable Multilinear Bases for Image Compression ​

Author: Shiwen An, Zhongyi Ni, Huanhai Zhou, Jin-Guo Liu
Published: 9/1/2026, 4:00:00 AM
Categories: eess.IV, cs.CV, cs.LG, math.OC, quant-ph

arXiv:2608.00053v3 Announce Type: replace-cross Abstract: The Discrete Fourier Transform (DFT), the Discrete Cosine Transform (DCT), and their block-wise variants underpin most deployed image and video codecs. Their effectiveness rests on three properties: their runtime is near-linear (up to a polyl...

📖 Read original article


579. Dynamically Allocating Evaluation Effort for Model Ranking ​

Author: Vil'em Zouhar, Julia Kreutzer, Alon Lavie, Tom Kocmi, Matt Post, Ond\v{r}ej Bojar, Mrinmaya Sachan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.03437v2 Announce Type: replace-cross Abstract: While human evaluation is the gold standard in many NLP tasks, it suffers from prohibitive costs and poor scalability. When identifying top-performing models, typical evaluation protocols waste effort by exhaustively evaluating all models on ...

📖 Read original article


580. SDF-Aware Weighting: Adaptive Eikonal Regularisation for Three-Dimensional Level-Set Physics-Informed Neural Networks ​

Author: Muhammad Akbar Khan
Published: 9/1/2026, 4:00:00 AM
Categories: physics.flu-dyn, cs.LG, physics.comp-ph

arXiv:2608.08322v2 Announce Type: replace-cross Abstract: Adaptive loss-balancing schemes for physics-informed neural networks rest on a premise that every residual should be driven to zero. For level-set advection with an eikonal regulariser that premise fails: the eikonal term penalises deviation ...

📖 Read original article


581. MITRE-SAGE: A Multi-Agent Cybersecurity Question-Answering Model ​

Author: Ali Habibzadeh, Farid Feyzi, Reza Ebrahimi Atani
Published: 9/1/2026, 4:00:00 AM
Categories: cs.IR, cs.CR, cs.LG, cs.MA

arXiv:2608.16921v3 Announce Type: replace-cross Abstract: Effective cybersecurity operations require timely and accurate analysis of large-scale heterogeneous security information; however, analysts increasingly struggle with information overload, alert fatigue, and time-constrained decision-making....

📖 Read original article


582. A Deterministic Constant-Competitive Algorithm for Dynamic Mixture-of-Experts Serving ​

Author: Ian D'Ambrosio (Nth Research Collective)
Published: 9/1/2026, 4:00:00 AM
Categories: cs.DS, cs.LG

arXiv:2608.16947v2 Announce Type: replace-cross Abstract: Dynamic Mixture-of-Experts Serving allocates k replica GPUs among m experts as workloads change. At each round, the online algorithm sees the current workload, chooses integral replica counts, and pays bottleneck service cost plus replica mov...

📖 Read original article


583. A Comprehensive Review of Large Language Models for Nanophotonics: From Surrogate Modeling to Autonomous Design ​

Author: Huanshu Zhang, Kegeng Tang, Lei Kang, Sawyer D. Campbell, Zihao Wang, Douglas H. Werner
Published: 9/1/2026, 4:00:00 AM
Categories: physics.optics, cs.LG

arXiv:2608.18279v2 Announce Type: replace-cross Abstract: Metasurfaces have revolutionized the development of photonic devices by enabling unprecedented precision in light manipulation. However, their design processes are often constrained by computationally expensive simulations and complex high-di...

📖 Read original article


584. DELE-w0.5: Inferring Action from Future Latent State for Robotic Manipulation ​

Author: Fenghao Lei, Zhixiong Huang, Long Yang, Jiabao Chen, Peilin Huang, Han Fu, Zhuo Li, Xiaoxue Ren
Published: 9/1/2026, 4:00:00 AM
Categories: cs.RO, cs.AI, cs.CV, cs.LG

arXiv:2608.22067v4 Announce Type: replace-cross Abstract: World-Action Models (WAMs) build robot control on video-generation backbones, which jointly predict dense future visual trajectories and robot actions. We argue that video generation is an unnecessary intermediate objective for world-action m...

📖 Read original article


585. GTA-RAG: Graph-Trajectory-Augmented Reinforcement Learning for Multi-Turn Retrieval-Augmented Reasoning ​

Author: Jun Chen, Yongchao Liu, Pengyu Qiu, Jiajun Zheng, Juelu Zhang, Yujie Zeng, Qin Zhang, Ziyue Qiao, Xiao Luo
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CL, cs.LG

arXiv:2608.22479v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) enables LLMs to access external knowledge for answering knowledge-intensive questions. For complex multi-hop questions, multi-turn retrieval-augmented reasoning extends RAG into an iterative process that r...

📖 Read original article


586. LUCAID: Agentic Multimodal AI for Lung Cancer Precision Pathology ​

Author: Marie-Lisa Eich, Kai Standvoss, Timo Milbich, Alexander M"ollers, Miriam H"agele, Philipp Anders, Lars Tharun, Hanna Kontradiuk, Sebastian Kons, Nader Aldoj, Recepcan Adig"uzel, Adam Narai, Lukas H"onig, Jonathan Striebel, Binru Yang, Mihnea P. Dragomir, Marvin Sextro, Philipp Keyl, Philipp Jurmeister, Rosemarie Krupar, Evelyn Ramberger, James Wells, Julika Ribbat-Idel, Andreas Kunft, Hussam Shuaib, Christian Groh'e, Reinhard B"uttner, David Horst, Klaus-Robert M"uller, Lukas Ruff, Maximilian Alber, Frederick Klauschen, Simon Schallenberg
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CV, cs.AI, cs.LG

arXiv:2608.23803v2 Announce Type: replace-cross Abstract: Lung cancer tissue diagnostics is complex, as therapy decisions in precision oncology rely on the integration of histomorphological, immunohistochemical, and molecular features. Yet pathological assessment remains largely visual and semi-quan...

📖 Read original article


587. SatDL: Jointly Optimizing Data Redistribution and Training for Satellite-Based Distributed Learning ​

Author: Hao Wu, Kin Whye Chew, Yizhan Han, Han Li, Jingxian Wang
Published: 9/1/2026, 4:00:00 AM
Categories: cs.DC, cs.LG

arXiv:2608.24516v2 Announce Type: replace-cross Abstract: Satellite-based distributed learning promises to train machine-learning models directly in orbit using massive, globally dispersed sensor data, thereby avoiding large-scale data downloads to ground servers. However, training convergence is si...

📖 Read original article


588. Forecasting Weather-Driven Price Dynamics Across Sri Lankan Tea Market Catalogues ​

Author: Hesandi Mallawarachchi, Senilka Madurapperumage, Nadil Kulathunge, Thilokya Angeesa, Nethsith Gunaweera, Sandeepa Weerasekara, Patalee Narasinghe, Nisansa de Silva, Sandareka Wickramanayake
Published: 9/1/2026, 4:00:00 AM
Categories: econ.GN, cs.LG, q-fin.EC

arXiv:2608.24894v2 Announce Type: replace-cross Abstract: The Colombo Tea Auction (CTA) plays a vital role in determining global tea prices, yet the relationship between local weather conditions and price behavior across different tea catalogues has not been thoroughly explored. In this study, we de...

📖 Read original article


589. Towards Reliable, Generalizable, and Specific In-Context Knowledge Editing via Multi-Objective Reinforcement Learning ​

Author: Xuzhong Wang, Maiqi Jiang, Tejal Nair, Girija Bhusal, Yanfu Zhang, Haipeng Chen
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.25100v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are powerful but limited by static parametric knowledge that becomes outdated once pretraining ends. Knowledge editing addresses this problem by updating model behavior on target facts without full retraining. In ...

📖 Read original article


590. Towards Large-Scale Heterogeneous Data Organization for Scientific Foundation Models: A Nuclear Fusion Case Study ​

Author: Nathaniel Chen, Kouroche Bouchiat, Peter Steiner, Azarakhsh Jalalvand, SangKyeun Kim, Egemen Kolemen
Published: 9/1/2026, 4:00:00 AM
Categories: physics.plasm-ph, cs.LG

arXiv:2608.27578v2 Announce Type: replace-cross Abstract: Training effective foundation models requires massive and organized datasets, yet scientific domains such as nuclear fusion present unique challenges due to largely heterogeneous and sparse data. Here we characterize the data used in developi...

📖 Read original article


591. RealSWE: A Compositional Evaluation of Coding Agents under Realistic User Requests ​

Author: Gyuhyeong Kim, Hyojung Gwon, Jeonghyeon Kim, Kyuhong Shim, Sunjae Lee
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG, cs.SE

arXiv:2608.27831v2 Announce Type: replace-cross Abstract: Coding agents are now commonly evaluated on the SWE-bench family of benchmarks, whose tasks are built from curated GitHub issues: long, structured, and information-rich. Real user requests, however, are typically far shorter and less structur...

📖 Read original article


592. Not to Break, but to Attest: Adversarial Probes for Privacy-Preserving LLM Verification ​

Author: Cameron Wilding, Mina Shaker, Fatemeh Ganji
Published: 9/1/2026, 4:00:00 AM
Categories: cs.CR, cs.AI, cs.LG

arXiv:2608.27954v2 Announce Type: replace-cross Abstract: Post-deployment changes to large language models can alter behavior while leaving routine outputs largely unchanged, creating a challenge for AI governance when model weights are proprietary. We present a privacy-preserving zk-SNARK-based aud...

📖 Read original article


593. Timing-Aware Repurchase Prediction for Web-Scale E-Commerce: Survival Models for Multi-Surface Grocery Recommendation ​

Author: Akshay Kekuda, Shreeranjani Srirangamsridharan, Ishan Bhatt, Yanan Cao, Sinduja Subramaniam, Evren Korpeoglu, Kaushiki Nag, Kannan Achan
Published: 9/1/2026, 4:00:00 AM
Categories: cs.AI, cs.LG

arXiv:2608.28393v2 Announce Type: replace-cross Abstract: Repurchase recommenders in e-commerce are commonly framed as a binary question asking "will this customer buy this item within W days", a formulation that requires a separately trained model for every horizon of interest. We replace this stac...

📖 Read original article