cs.LG · 2026-07-16 · No. 55

Machine Learning, 2026-07-16.

66 new papers in cs.LG. Titles, authors, abstracts. Links to arXiv. Want this in your inbox every morning? Subscribe →

01 — The papers

66 entries
  1. 01

    Leveraging unlabelled data for generalizable neural population decoding

    Ximeng Mao, Nanda H. Krishna, Avery Hee-Woon Ryoo, Matthew G. Perich, Guillaume Lajoie

    cs.LG · q-bio.NC

    Robust and accurate neural decoders are integral to neurotechnologies such as brain-computer interfaces and closed-loop experiments. Recent work has shown that tokenizing neural data at the spike level facilitates multi-session pretraining and delivers state-of-the-art decoding performance. However, current spike-based models are restricted to supervised learning (SL), limiting training to datasets with paired behavioural labels. To address...

    arxiv.org/abs/2607.14086 · PDF

  2. 02

    Linear Independent Component Analysis via Optimal Transport

    Ashutosh Jha, Michel Besserve, Simon Buchholz

    cs.LG · stat.ML

    Linear Independent Component Analysis (ICA) recovers jointly independent source signals from their linear mixtures. To achieve this, classical ICA algorithms attempt to maximize non-Gaussianity, measured by negentropy, which is linked to independence by information theory. Because exact negentropy optimization is intractable, they rely on proxy contrast functions, such as fourth-order cumulants, and parametric log-likelihoods. We propose...

    arxiv.org/abs/2607.14081 · PDF

  3. 03

    MetaPerch: Learning from metadata for bioacoustics foundation models

    Mustafa Chasmai, Vincent Dumoulin, Jenny Hamer

    cs.LG · cs.SD

    Bioacoustic foundation models rely on large-scale citizen science platforms like Xeno-Canto for geographically and ecologically diverse data. Recent work has shown that supervision alone can produce SotA species detection models when trained on this large-scale data -- however, there remains unutilized potential in the form of recording metadata readily available within these community-driven data hubs. In this work, we explore the use of...

    arxiv.org/abs/2607.14072 · PDF

  4. 04

    Improving Wind and Solar Power Prediction with Efficient Wrapper-based Feature Selection: An Empirical Study

    Daniel Grillmeyer, Marius Hadry, Michael Stenger, Vanessa Borst, Veronika Lesch, Samuel Kounev

    cs.LG · cs.AI

    With rising global energy demand and growing awareness of climate change and its impacts, the share of renewable energies in the global energy mix continues to grow. Unlike conventional power generation, the output of renewable energy sources cannot be controlled as consistently due to their dependence on environmental conditions. Therefore, reliable prediction of current and future energy production is essential. In this paper, we report...

    arxiv.org/abs/2607.14024 · PDF

  5. 05

    Transforming Rank: How Architecture Navigates the Spectral Pathologies of Depth

    Katie Everett

    cs.LG · cs.AI

    We investigate how each component of the Transformer feedforward block architecture design determines how much rank survives across depth at initialization. We reinterpret skip connections and normalization, long understood as controlling magnitude, as mechanisms for preserving gradient rank across depth, since the very matrix multiplications and nonlinear activations that make the network expressive also reduce the rank. We show that skip...

    arxiv.org/abs/2607.14018 · PDF

  6. 06

    Lighthouse RL: Sample-Efficient Circuit Optimization via Strategic Reset Points

    Mustafa Emre Gürsoy, Stefan Uhlich, Ryoga Matsuo, Yağız Gençer, Arun Venkitaraman, Chia-Yu Hsieh, Andrea Bonetti,...

    cs.LG · cs.AR

    In this paper, we introduce Lighthouse RL, a sample-efficient reinforcement learning (RL) approach for analog circuit sizing. Traditional methods lack generalization across different performance targets, while standard RL approaches waste resources exploring unpromising regions. Our method addresses these inefficiencies through a strategic reset strategy that initializes episodes from high-performing configurations discovered during training,...

    arxiv.org/abs/2607.14008 · PDF

  7. 07

    Lyapunov Exponent as Physics-Informed Dense Reward: RL Discovery of Stabilization Beyond the Kapitza Pendulum

    Slava Andrejev

    cs.LG

    We suggest using the Lyapunov characteristic exponent (LCE) as a dense reward signal for the reinforcement learning problem of stabilizing the inverted pendulum with vertical motion. With LCE, the agent not only successfully found the oscillatory motion known as the Kapitza pendulum but also damped the pendulum's pivoting, leaving it in a strictly upright position.

    arxiv.org/abs/2607.14001 · PDF

  8. 08

    TRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon Agents

    Leitian Tao, Baolin Peng, Wenlin Yao, Tao Ge, Hao Cheng, Mike Hang Wang, Jianfeng Gao, Sharon Li

    cs.LG

    Multi-turn agents solve complex tasks through extended sequences of tool interactions before producing a final answer, making credit assignment a fundamental challenge during post-training. Outcome rewards provide reliable supervision for short-horizon reasoning, but become sparse and high-variance as trajectories grow to tens or hundreds of tool calls. They can also be misleading: a failed rollout may contain many useful actions that move...

    arxiv.org/abs/2607.13988 · PDF

  9. 09

    VAIOM: Continuous-Input, Discrete-Output Decoder-Only Financial Sequence Modeling

    Yiming Ma, Xinyu Chen

    cs.LG · q-fin.CP

    Financial observations are continuous, heterogeneous, and noisy, whereas decoder-only next-token models are usually built around discrete symbolic inputs. We introduce Vector-Input Autoregressive Inference for Ordinal-Return Modeling (VAIOM), a decoder-only Transformer for probabilistic next-return modeling on one-hour foreign-exchange bars. VAIOM separates input representation from output likelihood: continuous multivariate financial-event...

    arxiv.org/abs/2607.13929 · PDF

  10. 10

    An Efficient Newton Algorithm for Nonnegative Matrix Factorization with the Kullback-Leibler Divergence

    Damien Lesens, Jérémy E. Cohen, Bora Uçar

    cs.LG

    Nonnegative Matrix Factorization (NMF) is a fundamental tool in unsupervised learning, which approximates a nonnegative matrix by the product of two low-rank nonnegative factors. The Kullback-Leibler (KL) divergence is best suited to measure the data to model discrepancy when the decomposed data sample follows a Poisson distribution, which is the case for count datasets such as term-document matrices or images. Most KL-NMF algorithms in the...

    arxiv.org/abs/2607.13919 · PDF

  11. 11

    RF Spectrogram Anomaly Detection with Quantum Kitchen Sinks: Architecture, Representation, and Hardware Validation

    Abdallah Aaraba, Alexis Vieloszynski, Remon Polus, Ola Ahmad, Soumaya Cherkaoui

    cs.LG

    The broadcast nature of wireless channels exposes radio-frequency (RF) networks to anomalous and malicious transmissions, making anomaly detection a fundamental requirement for secure spectrum management. Quantum Kitchen Sinks (QKS) offer a lightweight hybrid quantum feature map suitable for near-term quantum devices, yet their behavior on structured signal data remains poorly understood. In this paper, we extend the standard QKS template...

    arxiv.org/abs/2607.13897 · PDF

  12. 12

    PiVoT: A Variational Solution for Real-time Large-scale Multi-object Detection and Tracking under Heavy Clutter

    Runze Gan, Qing Li, Simon J. Godsill, Mike E. Davies, James R. Hopgood

    cs.LG · cs.CV · eess.SP

    Multi-object detection and tracking from noisy point clouds remain challenging in many data-scarce radar applications. Current Bayesian trackers based on Poisson measurement models offer a training-free solution but struggle to achieve accuracy and efficiency under severe clutter, large object populations, and full-resolution Doppler point clouds. We address this with PiVoT, a fast, clutter-resilient multi-object tracker for both positional...

    arxiv.org/abs/2607.13891 · PDF

  13. 13

    Task-Oriented Sensing and Covert Transmissions for Collaborative Multi-AUV Systems

    Xueyao Zhang, Chenyang Yan, Bo Yang, Xuelin Cao, Zhiwen Yu, Bin Guo, George C. Alexandropoulos, Merouane Debbah, Chau Yuen

    cs.LG

    In underwater covert cooperative missions, autonomous underwater vehicles (AUVs) often cannot rely on active sonar to continuously obtain complete information, since active sensing and frequent communications increase the risk of exposure. As a result, AUVs primarily rely on passive observation, an approach that yields incomplete local perception and limited task efficiency. Although underwater acoustic communications can mitigate this...

    arxiv.org/abs/2607.13880 · PDF

  14. 14

    AI-Augmented Adaptive Digital Twin Modeling for Brain Tumor Evolution Prediction and Treatment Scheduling

    Wenxi Liu, Michael Trimboli, Xianqi Li

    cs.LG

    Brain tumor progression exhibits spatially heterogeneous growth, patient-specific treatment response, and complex interactions with surrounding anatomy, making accurate long-term prediction challenging. We propose an AI-augmented adaptive digital twin (DT) framework for brain tumor evolution prediction and treatment scheduling. The framework integrates an interpretable reaction--diffusion (RD) model, a 3D residual learning module for...

    arxiv.org/abs/2607.13877 · PDF

  15. 15

    Relevance-Aware Rule: Structural Deletion of Irrelevant Conditions in Decision Trees

    Jung-Sik Hong, Jeongeon Lee, Min Kyu Sim, Sangheum Hwang

    cs.LG

    Decision trees generate interpretable if--then rules, yet they contain irrelevant conditions (IRCs). These IRCs arise from the structural mechanism of tree splitting and persist even in modern optimal sparse tree induction algorithms. Existing IRC deletion methods overlook this structural mechanism; therefore, they either preserve the original tree too loosely to remain reliable, or too strictly to achieve meaningful simplification. This...

    arxiv.org/abs/2607.13874 · PDF

  16. 16

    Heavy-Tailed Flow Matching via Random Clocks

    Zhouhao Yang, Yezhen Wang, Kenji Kawaguchi, Vladimir Braverman, Haoyang Cao

    cs.LG · stat.ML

    Heavy-tailed data arise in many domains where rare events carry disproportionate importance, such as imbalanced image datasets, financial returns, and weather extremes. Standard diffusion and flow-matching models typically begin from Gaussian noise or Gaussian source distributions, which yield tractable training targets but provide a poor inductive match for heavy-tailed data. We propose Heavy-Tailed Flow Matching via Random Clocks (HTFM), a...

    arxiv.org/abs/2607.13841 · PDF

  17. 17

    NodeImport: Imbalanced Node Classification with Node Importance Assessment

    Nan Chen, Zemin Liu, Bryan Hooi, Bingsheng He, Jun Hu, Jia Chen

    cs.LG · cs.AI

    In real-world applications, node classification on graphs often faces the challenge of class imbalance, where majority classes dominate training, resulting in biased model performance. Traditional GNNs often struggle in such scenarios, as they tend to overfit to majority classes while underrepresenting minority classes. Existing solutions, which either prioritize nodes based on class size or synthesize new nodes for minority classes, often...

    arxiv.org/abs/2607.13837 · PDF

  18. 18

    Mono-Z Dark Matter Search with Neural Spline Flows Using CMS Run 2015D Open Data

    Hitesh Rasineni, Bhavishya Chebrolu

    cs.LG · hep-ex

    We report a search for dark matter (DM) produced in association with a leptonically decaying \(Z\) boson at \(\sqrt{s}=13\) TeV using CMS Run 2015D open data corresponding to an integrated luminosity of \(2.32\,\mathrm{fb}^{-1}\) together with simplified-model Monte Carlo simulation. Events are selected in the mono-\(Z\rightarrow\ell^+\ell^-\) final state in both the \(μμ\) and \(ee\) channels. Forty kinematic observables are extracted from...

    arxiv.org/abs/2607.13771 · PDF

  19. 19

    MxGPS: Multiplex Graph Transformers for a Power Grid Foundation Model

    Charilaos Papaioannou, Ioannis Tsantilas, Dimitris Giannakakos, Vasilis Michalakopoulos, Sotiris Pelekis, Vangelis...

    cs.LG · cs.AI

    Single-task fine-tuning of graph neural networks (GNNs) for power grid problems exhibits a systematic failure mode: models that achieve the lowest in-distribution error degrade the most under topology shift. We term this topology overfitting: the tendency of task-specific gradient signals to encode relational structure particular to the training topologies rather than the underlying physics, causing models to fail on unseen grids despite...

    arxiv.org/abs/2607.13763 · PDF

  20. 20

    Algebraic Representability as the Limiting Regime of Grokking: An Exactly Solvable Model with Holomorphic Activations

    Chon-Fai Kam, Xavier Cadet, Miloud Bessafi, Frederic Cadet

    cs.LG · stat.ML

    Neural networks trained on modular arithmetic exhibit grokking, a delayed transition from memorisation to generalisation known to depend on model capacity: too little and the network memorises slowly or not at all, too much and it generalises almost immediately. What happens at the extreme of this spectrum, when the architecture's expressible function class collapses to a finite-dimensional algebraic variety? We study two-layer networks with...

    arxiv.org/abs/2607.13749 · PDF

  21. 21

    Implementations of Quantum and Classical Topology-Aligned Architectures for Molecular Property Prediction

    James T. Pegg, Hubert Okadome Valencia, Ronin Wu

    cs.LG

    For low-data and resource-constrained regimes typical of quantum chemistry, parameter-efficient learning is a key objective. Here, we propose a topology-aligned inductive bias in which the model architecture mirrors the molecular bond graph: atoms map to a fixed register of computational units, and bonds determine which pairs interact through shared learnable parameters. This principle is instantiated in two architectures: a variational...

    arxiv.org/abs/2607.13737 · PDF

  22. 22

    Constraint-Driven Model Optimization: An Industry Framework for Selecting Compression and Acceleration Techniques in Modern Machine Learning Systems

    Dhruv Shivkant, Saket Mohanty, Utkarsh Wadhwa

    cs.LG

    The rapid deployment of machine learning systems across cloud, edge, and enterprise environments has brought model optimization to the forefront of systems-engineering. Despite a rich literature spanning quantization, pruning, knowledge distillation, parameter-efficient fine-tuning (PEFT), and inference-time optimization, practitioners are often left navigating these techniques through heuristics rather than principled methodology. We argue...

    arxiv.org/abs/2607.13735 · PDF

  23. 23

    DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention

    Xing Lei, Wenyan Yang, Xuetao Zhang, Donglin Wang

    cs.LG · stat.ML

    Goal-conditioned reinforcement learning hinges on how the goal is encoded. Contrastive, metric, temporal-distance, and information-theoretic encoders differ in objective. They still share one trait. None of them sees the current state. Such a state-independent embedding cannot mark which part of the goal still needs action. The policy must then recover that cue by inverting both encoders. We propose DAGR. It refines the static embedding of...

    arxiv.org/abs/2607.13731 · PDF

  24. 24

    Conditional Invertible Neural Networks for Data-Driven UAV Control: A 2-D Proof of Concept

    Christian Wittke, Stephan Myschik, Oliver Niggemann

    cs.LG · eess.SY

    We investigate conditional invertible neural networks (cINNs) as probabilistic inverse-dynamics models for multirotor control. For a planar X8 coaxial multicopter, we learn $p(u \mid s_t, c_t)$ from an incremental nonlinear dynamic inversion (INDI) teacher using rational-quadratic spline coupling and invertible linear mixing. Open-loop reproduction reaches $R^2 = 0.944$, mean CRPS 0.0915, and log-probability-error correlation $ρ= -0.60$. Over...

    arxiv.org/abs/2607.13703 · PDF

  25. 25

    Microstructure-Conditioned Surrogate Models for Graded Multiscale Optimization of Mycelium Composites

    J. Storm, I. B. C. M. Rocha, S. Schyck, K. Masania, F. P. van der Meer

    cs.LG · cond-mat.mtrl-sci · math.NA · physics.comp-ph

    Emerging sustainable materials increasingly rely on engineered hierarchy and microstructure to achieve control of their properties and mechanical behavior. Optimizing these materials with controllable microstructures requires efficient multiscale simulations. Data-driven surrogate models for the microscale can accelerate multiscale simulations, but require large amounts of data even for a fixed microstructure. When a range of microstructures...

    arxiv.org/abs/2607.13688 · PDF

  26. 26

    Optimal and Efficient Contextual Combinatorial Semi-bandits with General Function Approximation

    Hao Qin, Chicheng Zhang

    cs.LG

    We study the contextual combinatorial semi-bandit (CCSB) problem with general reward function approximation. At each round, the learner observes a context, selects a combinatorial action consisting of a subset of basic arms, and receives the reward of each selected arm; the goal is to maximize the cumulative reward over time. We propose SquareCB.Comb, a computationally efficient algorithm that, at each round, solves a convex optimization...

    arxiv.org/abs/2607.13686 · PDF

  27. 27

    The Hyperspherical Geometry of CLIP Latent Space: A Semantic Mixture Model

    Zijie Yu, Gaowen Liu, Ramana Rao Kompella, Philip S. Yu, Yue Song

    cs.LG

    Contrastive Language-Image Pretraining (CLIP) representations form a semantic embedding space governed by cosine similarity, reflecting an intrinsic hyperspherical geometry. However, existing probabilistic interpretations typically rely on Gaussian assumptions, which fail to capture this directional and multimodal structure. We propose a principled density model for the CLIP latent space based on Mixtures of von Mises-Fisher (MovMF)...

    arxiv.org/abs/2607.13660 · PDF

  28. 28

    Maximally Robust Satisficing Bayesian Optimization

    Samuli Kinnunen, Petrus Mikkola, Antti Niskanen, Arto Klami

    cs.LG

    Many design tasks can be cast as black-box function optimization, enabling use of Bayesian optimization to find an ideal design with minimal number of trials. However, often we do not actually need the optimum but instead a sufficiently good solution is enough, for instance a material that is durable enough for its intended use. In most cases there are multiple satisfactory solutions, forming a superlevel set of the function, raising a key...

    arxiv.org/abs/2607.13652 · PDF

  29. 29

    Consensus as Privileged Context for Label-Free Self-Distillation

    John Gkountouras, Josip Jukić, Ivan Titov

    cs.LG · cs.AI · cs.CL

    Sampling multiple solutions and returning the majority answer is among the most reliable ways to improve the reasoning accuracy of large language models without labels, and a growing family of methods converts this consensus signal into training supervision. However, existing approaches use consensus only in restricted forms: as a filter that selects solutions for fine-tuning, as a preference between answers, or as a scalar reward for...

    arxiv.org/abs/2607.13643 · PDF

  30. 30

    How the Hessian-Spectrum of Neural Networks Depends on Data

    Jasraj Singh, Enea Monzio Compagnoni, Antonio Orvieto

    cs.LG

    The Hessian matrix is an important quantity of interest when it comes to studying the loss landscape and optimization dynamics in deep learning, as well as designing measures of generalization, second-order learning algorithms, etc. Prior works have focused on empirical results or pursued a theoretical treatment under overly simplified settings. In this work, we derive the eigenvalues of the Hessian of linear networks with arbitrary widths...

    arxiv.org/abs/2607.13631 · PDF

  31. 31

    FastCentNN: Accelerating Centroid Neural Network with Entropy Proxy

    Le-Anh Tran

    cs.LG · cs.CV

    Centroid neural network (CentNN) is an unsupervised competitive learning algorithm in which centroid splitting is triggered only after strict local stabilization, often leading to prolonged low-movement training phases before model expansion. This report proposes FastCentNN, an accelerated variant that addresses this inefficiency by introducing an early splitting strategy based on the total centroid movement per epoch, which serves as a...

    arxiv.org/abs/2607.13613 · PDF

  32. 32

    The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models

    Fabio Arnez, Alexandra Gomez-Villa

    cs.LG · cs.AI

    Joint-Embedding Predictive Architectures (JEPAs) are the dominant design for latent world models, yet they are usually justified by empirical performance rather than a normative principle. We show that the choice of anti-collapse regulariser determines whether a JEPA's training objective, a prediction loss plus a weighted embedding regulariser, is a valid Active Inference (AIF) variational free energy. We organise four non-contrastive...

    arxiv.org/abs/2607.13612 · PDF

  33. 33

    Gauge-Invariant, Parameter-Insensitive Regularization for Potential Recovery from Flow on Directed Graphs

    Mohammad Forouhesh

    cs.LG · cs.IR · eess.SP · stat.ML

    Recovering a latent potential from observed flow on a directed graph (a discrete Poisson problem with Dirichlet boundaries) is ill-posed, and the standard fix backfires: ridge regularization shrinks toward a gauge-meaningless origin, collapsing and reversing the recovered ordering ($+0.81\to-0.42$ rank correlation against a planted ground truth). The gauge-invariant graph Dirichlet energy removes the hazard and delivers...

    arxiv.org/abs/2607.13609 · PDF

  34. 34

    Structured Reinforcement Learning for Bayesian Persuasion : Application to Intelligent Interactive Driving

    Merlin Paul, Anup Aprem

    cs.LG · eess.SP

    Interactive driving, wherein an intelligent lead vehicle equipped with real-time traffic data coordinates route choices of connected vehicles, offers a promising approach to dynamic traffic management. To address the challenge of harmonising decisions, this paper considers the strategic information revealing framework of Bayesian persuasion. Here, the principal (lead vehicle) aims to guide the agent's (connected vehicle) partially observable...

    arxiv.org/abs/2607.13576 · PDF

  35. 35

    From Novice to Expert: Cost-Aware Bandits for Evolving Worker Performance in Crowdsensing

    Yin Huang, Qingsong Liu, Jie Xu

    cs.LG

    Mobile crowdsensing (MC) recruits mobile users to perform sensing tasks using their smartphones, enabling large-scale applications such as traffic monitoring and environmental sensing. A fundamental challenge is online worker recruitment under uncertainty, where the platform must learn workers' sensing performance while operating with a limited budget. Existing learning-based MC recruitment methods typically assume that each worker's sensing...

    arxiv.org/abs/2607.13546 · PDF

  36. 36

    Clustering algorithms for multivariate wind farm SCADA data filtering

    Nicolò Italiano, Vasilis Pettas, Tuhfe Göçmen, Nicolaos A. Cutululis

    cs.LG

    During wind farm operation, Supervisory Control and Data Acquisition (SCADA) systems record numerous anomalies, transients, and specific operational modes, leading to large datasets. However, for a wide range of applications, only measurements corresponding to normal operation are required and, therefore, the SCADA data must be filtered. For this purpose, several methods have been proposed to automate and replace manual filtering conducted by...

    arxiv.org/abs/2607.13544 · PDF

  37. 37

    ExTernD: Expanded-Rank Ternary Decomposition Ternary LLM PTQ with Accuracy Approaching Any Quantization Level

    Chethan Reddy G. P

    cs.LG · cs.AI

    We introduce ExTernD (Expanded-rank Ternary Decomposition), a post-training factorization of each LLM weight matrix $A \in \mathbb{R}^{m \times n}$ into $A \approx B \mathrm{diag}(D) C$ with ternary factors $B \in \{-1,0,+1\}^{m \times k}$, $C \in \{-1,0,+1\}^{k \times n}$ and a real scale vector $D \in \mathbb{R}^k$. The inner rank $k = μ\min(m,n)$ is deliberately expanded beyond full rank ($μ> 1$), so that components past full rank correct...

    arxiv.org/abs/2607.13511 · PDF

  38. 38

    CDS: Counterfactual Directionality Score for Structured Interventions in Spatial Graphs

    Humaira Anzum, Md Ishtyaq Mahmud, Jagan Mohan Reddy Dwarampudi, Tania Banerjee

    cs.LG

    Quantifying directional influence between node populations is a fundamental problem in graph-based modeling, particularly in spatial biological systems where cell-cell interactions shape functional outcomes. Existing approaches based on attention, attribution, or correlation capture associations but do not provide a principled framework for evaluating directional effects under controlled perturbations. We introduce a framework for structured...

    arxiv.org/abs/2607.13508 · PDF

  39. 39

    Factorized Spectral Representations for Reinforcement Learning

    Junyi Wu, Dan Li

    cs.LG

    Learning a compact model of the world from interaction data is central to sample-efficient deep reinforcement learning. Spectral representation methods have become the leading paradigm for representation learning in continuous control by taking a matrix view of the transition kernel, with state-action pairs on one side and next states on the other, and learning a low-rank factorization through self-supervised contrastive objectives. We take...

    arxiv.org/abs/2607.13498 · PDF

  40. 40

    A VAE-Driven Multi-Task Satellite-Aided Semantic Communication Framework for 6G-Enabled Connected Autonomous Vehicles

    S. M. Abtahiul Alam, Niloy Das, Apurba Adhikary, Yu Qiao, Zhu Han, Choong Seon Hong

    cs.LG · cs.IT

    The development of smart transportation systems and the introduction of 6G wireless communication technologies have significantly changed vehicle network topologies. Future connected autonomous vehicle (CAV) networks require bandwidth-efficient, reliable, and low-latency communication for safety-critical applications such as traffic sign recognition and decision-making. Conventional communication systems transmit raw data regardless of task...

    arxiv.org/abs/2607.13494 · PDF

  41. 41

    DeepLoop: Depth Scaling for Looped Transformers

    Shuzhen Li, Yifan Zhang, Jiacheng Guo, Quanquan Gu, Mengdi Wang

    cs.LG · cs.AI

    Looped Transformers scale sequential computation by applying a compact stack of physical blocks for multiple rounds, increasing unrolled depth without increasing stored parameters. This reuse changes the residual-scaling problem: in an untied Transformer, each residual branch receives and applies its own parameter update, whereas in a looped Transformer one shared update aggregates gradients from repeated visits and is read back by those same...

    arxiv.org/abs/2607.13491 · PDF

  42. 42

    Explainable Artificial Intelligence for Anomaly Detection in Banking Transactions: An Internal Audit Perspective

    Anupa Lodhi

    cs.LG · cs.AI

    The banking sector increasingly relies on automated systems to monitor electronic transactions for signs of fraud, yet conventional rule-based approaches struggle with high false-positive rates and offer no justification for their outputs, limiting their utility for compliance teams. This paper introduces an Explainable Artificial Intelligence (XAI) framework tailored for banking transaction anomaly detection within internal audit workflows....

    arxiv.org/abs/2607.13469 · PDF

  43. 43

    PQFA: Parallel Quantum Feature Augmentation of Fused Representations for Multimodal Classification

    Mingzhu Wang, Yun Shang

    cs.LG · quant-ph

    Most multimodal learning methods improve how heterogeneous representations are aligned and fused, while post-fusion enhancement remains less explored. We propose Parallel Quantum Feature Augmentation (PQFA), a hybrid quantum-classical framework that applies multiple shallow variational quantum circuits to fused multimodal features. Text and image representations extracted by frozen RoBERTa and ViT encoders are processed through bidirectional...

    arxiv.org/abs/2607.13466 · PDF

  44. 44

    Distributionally Robust and Safe Imitation Learning

    Ahmed Aboudonia, Naira Hovakimyan

    cs.LG · eess.SY

    Imitation learning (IL) has achieved remarkable success in complex decision-making tasks. However, its performance is highly sensitive to distribution shifts, which can pose significant safety risks. We propose a distributionally robust and safe IL framework that explicitly addresses both policy-induced and uncertainty-induced distribution shifts. Our approach develops a unified framework leveraging Taylor Series Imitation Learning (TaSIL) to...

    arxiv.org/abs/2607.13436 · PDF

  45. 45

    Local Redundancy: An Information-Theoretic Measure of Plasticity from Synthetic Memorization

    Jiaxuan Cheng

    cs.LG

    Plasticity -- a neural network's ability to adapt to new tasks -- is critical for continual and transfer learning. Existing measures, such as effective rank, dead neuron fraction, and weight norm, lack theoretical grounding and correlate poorly with performance on new tasks. We introduce local redundancy, an information-theoretic measure derived from universal compression theory. We define local redundancy as the worst-case redundancy of a...

    arxiv.org/abs/2607.13432 · PDF

  46. 46

    Discrete Diffusion Models: A Unified Framework from Tokenization to Generation

    Ye Yuan, Weien Li, Rui Song, Zeyu Li, Haochen Liu, Xiangyu Kong, Zixuan Dong, Linfeng Du, Zipeng Sun, Weixu Zhang,...

    cs.LG · cs.AI · cs.CL

    Discrete denoising diffusion models (DDMs) have recently emerged as a compelling alternative to autoregressive (AR) modeling for discrete data, offering parallel generation and iterative global refinement capabilities. Unlike continuous diffusion, where the state space is fixed, DDMs are fundamentally shaped by how the discrete state space is constructed: the tokenization scheme, the vocabulary topology, and domain-specific structural...

    arxiv.org/abs/2607.13431 · PDF

  47. 47

    PUe: Biased Positive-Unlabeled Learning Enhancement by Causal Inference

    Xutao Wang, Hanting Chen, Tianyu Guo, Yunhe Wang

    cs.LG

    Positive-Unlabeled (PU) learning aims to achieve high-accuracy binary classification with limited labeled positive examples and numerous unlabeled ones. Existing cost-sensitive-based methods often rely on strong assumptions that examples with an observed positive label were selected entirely at random. In fact, the uneven distribution of labels is prevalent in real-world PU problems, indicating that most actual positive and unlabeled data are...

    arxiv.org/abs/2607.13428 · PDF

  48. 48

    Data-Efficient Adaptation of LLMs via Attention Head Reweighting

    Tuomas Oikarinen, Zixiao Chen, Charlotte Siska, Tsui-Wei Weng, Chandan Singh, Jianfeng Gao

    cs.LG · cs.AI · cs.CL

    Learning effectively from limited data is critical in domains like security where labeled examples are scarce. Large language models (LLMs) have demonstrated some capabilities for data-efficient learning, especially through parameter-efficient adaptation methods, but continue to struggle when faced with few samples for difficult tasks. To meet this challenge, we propose Attention Head Reweighting (AHR), a data-efficient method that adapts...

    arxiv.org/abs/2607.13425 · PDF

  49. 49

    Temperature Scaling Is Not Enough: Calibration Gaps Under Human Label Distributions

    Wisdom Dogah

    cs.LG

    Temperature scaling is the dominant post-hoc calibration method in modern deep learning. Its theoretical justification rests on an assumption that is rarely stated explicitly: that ground-truth labels are one-hot and deterministic. In practice, labels are frequently soft, crowd-sourced, or genuinely distributional, reflecting real disagreement among human annotators rather than annotation noise. We study whether temperature scaling retains...

    arxiv.org/abs/2607.13423 · PDF

  50. 50

    OrDA: Orthogonal Disentanglement of Access Habits Framework for Homepage Marketing Block Recommendations

    Lingxiao Zhang, Xiaobo Li, Tao Xu

    cs.LG

    Clicks on homepage marketing blocks are driven by a dual-mechanism of content interest and access habits. However, habitual clicks often create Pseudo-Positives in marketing slots, where position advantage masks mediocre content quality, leading to biased recommendation ecosystems. We propose a framework called Orthogonal Disentanglement of Access habits (OrDA) to purify interest signals. OrDA utilizes a dual-tower structure with a gated...

    arxiv.org/abs/2607.13420 · PDF

  51. 51

    EXPLORE: Exploration with Guided Search for Analog Topology Generation using Language Models

    Guanglei Zhou, Chen-Chia Chang, Yikang Shen, Jonathan Ku, Isaac Jacobson, Jingyu Pan, Yiran Chen, Xin Zhang

    cs.LG

    Automating analog circuit topology design is essential to reduce the extensive manual effort required to meet increasingly diverse and customized application demands. Recent advances have applied sequence-to-sequence fine-tuning on pretrained language models to directly generate circuit topologies from user specifications in a single pass. However, these one-shot generation methods failed to generate complex circuits due to their...

    arxiv.org/abs/2607.13416 · PDF

  52. 52

    Is the Statistical Advantage Worth the Cost? An Empirical Comparison of KANs and MLPs for Structured Data Classification

    Matthew Steven P. Toledo, Justine Raphael H. Jacinto, Vivekjeet Singh Chambal, Rodolfo C. Camaclang, Jamlech Iram N....

    cs.LG · cs.AI

    This study presents an empirical benchmarking comparison between Kolmogorov-Arnold Networks (KANs) and Multi-Layer Perceptrons (MLPs) on structured tabular classification tasks. Motivated by the growing interest in KANs as an alternative function-approximating architecture, we evaluate their out-of-the-box performance on twelve publicly available datasets spanning binary, multiclass, multilabel, and ordinal problems. Both models were trained...

    arxiv.org/abs/2607.13413 · PDF

  53. 53

    Self-Improving is Often Sudden: Enlightenment-style Finetuning for Large-Scale Models

    Jing-Xiao Liao, Tianwei Zhang, Yu-Hao Jiang, Feifei Zhang, Hang-Cheng Dong, Feng-Lei Fan

    cs.LG

    The pursuit of autonomously self-improving models has attracted growing interest in the era of large-scale foundation models. Drawing inspiration from the concept of "enlightenment" or "aha moment" in human brain, we hypothesize that large models exhibit an analogous enlightenment phenomenon-a latent capacity for sudden capability boost. Then, we propose Enlightenment, a novel training-free post-tuning paradigm for large-scale models. Our...

    arxiv.org/abs/2607.13395 · PDF

  54. 54

    Where Should RL Post-Training Compute Go? Model Size, Search, Learning, and Feedback

    Patrick Wilhelm, Odej Kao

    cs.LG · cs.CL

    Reinforcement Learning (RL) post-training is increasingly used to adapt foundation models for reasoning, planning, and feedback-driven robot-learning pipelines, but constrained post-training resources are often summarized by a single total FLOP budget. We study the fixed-budget decision problem behind this practice: under the same post-training budget, should one use a larger policy, train a smaller policy longer, generate more rollout...

    arxiv.org/abs/2607.13389 · PDF

  55. 55

    Weight Feedback Computes the Jacobian Transpose Locally in Modern Deep Networks

    Junlong Shen, Xingyu Li

    cs.LG

    Predictive Coding (PC) offers a biologically motivated alternative to backpropagation via local weight updates, yet routing error between layers still relies on an autograd Jacobian-transpose ($J^\top$) product - the last non-local operation in PC. We show that this dependency is largely avoidable. For any layer $f(x)=\mathrm{Act}(\mathrm{Norm}(L(x)))$ with frozen normalization statistics, the exact $J^\top$ factors into three locally...

    arxiv.org/abs/2607.13380 · PDF

  56. 56

    Agora: Collective and Permissionless Internet-Scale Pretraining of Large Language Models

    Gil Avraham, Violetta Shevchenko, Hadi Mohaghegh Dolatabadi, Karol Pajak, James Snewin, Harry Xi, Rodney O'Donnell,...

    cs.LG · cs.DC

    Training large language models at the multi-billion to trillion parameter scale is confined to datacenters, where data-parallel (DP) and model-parallel (MP) techniques presume homogeneous accelerators, high-speed interconnects, and a single orchestrating entity. Frontier model development is thereby concentrated among the few groups able to assemble such clusters. Meanwhile, an enormous pool of compute remains unusable for training: consumer...

    arxiv.org/abs/2607.13332 · PDF

  57. 57

    Accuracy-Preserving Stability Regularization for Large-Scale Retail Demand Forecasting

    Jize Li, Jiani He, Dishu Yang, Dingyan Shang, Jingjing Liu, Shiqi Huang

    cs.LG

    Retail demand forecasts are reused across replenishment, capacity, labor, and transportation planning cycles. Point-error objectives do not constrain abrupt movement between adjacent forecasts, while post-hoc smoothing acts only after model fitting. We ask whether a training-time penalty on consecutive within-series movement can improve horizontal forecast-path stability without materially changing point accuracy. The penalty is evaluated in...

    arxiv.org/abs/2607.13331 · PDF

  58. 58

    Tabular Foundation Models for Discrete Choice Estimation

    Liu Liu, Dan Zhang

    cs.LG · cs.AI · econ.EM

    Tabular foundation models (TFMs) generate predictions on structured data via in-context learning, without task-specific estimation. We ask whether TFMs can be effectively applied to discrete choice, a central demand estimation framework in marketing and operations, and find that directly applying TFMs yields limited performance. The gap is structural: TFMs assume row-independent observations, whereas discrete choice is inherently set-valued...

    arxiv.org/abs/2607.13314 · PDF

  59. 59

    Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners

    Haseeb Shah, Lingwei Zhu, Adam White, Martha White

    cs.LG · cs.AI

    Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discovery and drinking water treatment, where reliability is essential and tuning budgets are limited. Actor-critic algorithms share a set of design decisions, such as how the policy is updated, how it represents the distribution over actions, how its gradient is estimated, and how often it is updated...

    arxiv.org/abs/2607.13274 · PDF

  60. 60

    Reassessing Muon for Matrix Factorization

    Ali Parviz, Gal Mishne, Alex Cloninger

    cs.LG · cs.AI

    Muon has recently emerged as a strong optimizer for large-scale deep learning, where it reshapes gradient updates through approximate orthogonalization and has been reported to outperform Adam and AdamW in large language model training. Its empirical success has motivated a growing body of theoretical work that interprets Muon as steepest descent under the spectral norm. Yet it remains unclear which of Muon's advantages stem from its update...

    arxiv.org/abs/2607.13246 · PDF

  61. 61

    EMAGN: Efficient Multi-Attention Graph Network via Learned Clustering for Scalable Traffic Forecasting

    Mingxing Xu, Rakesh Chowdary Machineni, Ke Liu, Xi Cheng, Chengqi Lu, Xin Hu, Lyuhao Chen, Xiangyu Li, Junwei You, Oliver Gao

    cs.LG · cs.AI

    Traffic forecasting is highly challenging due to complex and nonlinear spatial and temporal dependencies. Self-attention mechanisms have been widely adopted to model dynamic and long-range dependencies, achieving state-of-the-art performance, but suffer from limited scalability due to quadratic computational and memory complexity. To address this, we propose an Efficient Multi-Attention Graph Network (EMAGN) that linearises the spatial...

    arxiv.org/abs/2607.13241 · PDF

  62. 62

    Concurrent Image Understanding and Generation: Self-Correcting Coupled Markov Jump Processes

    Minh-Quan Le, Armand Comas, Alexandros Lattas, Stylianos Moschoglou, Pedro Vélez, Amit Raj, Aaron Germuth, Thabo...

    cs.LG

    Human cognition does not separate understanding and generation. A teacher at a whiteboard speaks and draws $\textit{together}$, each modality reshapes the other. In this paper, we bring this coupled loop to artificial systems. Masked Diffusion Models (MDMs) are ideally suited to this task, yet existing samplers either decode text and image interleavedly or independently update them in parallel branches that share only previous-step history,...

    arxiv.org/abs/2607.13188 · PDF

  63. 63

    SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy

    Yassine Chemingui, Chenhua Fan, Honghao Wei, Janardhan Rao Doppa

    cs.LG · cs.AI

    Safe reinforcement learning typically enforces safety by bounding expected cumulative costs, a criterion that often fails to detect rare but catastrophic tail events. To overcome these limitations, this paper introduces SteinGate, a boundary-aware distributional safety certificate that replaces fragile tail fitting with a robust consistency check using Kernelized Stein Discrepancy while accounting for boundary atoms induced by clipped costs....

    arxiv.org/abs/2607.13175 · PDF

  64. 64

    HEDGEHOG: Hierarchical Evaluation of Drug Generators Through Rigorous Filtration

    Daria A. Ryabchenko, Pavel Gurevich, Shamil Kadyrov, Daria Frolova, Kseniia Fedisheva, Sergei A. Nikolenko,...

    cs.LG · cs.SE

    Generative molecular models can support early drug discovery by proposing new candidate compounds de novo. In practice, useful candidates must balance target-relevant activity, synthetic accessibility, physicochemical properties, and other multiparameter design constraints. However, metrics commonly used to evaluate molecular generators only weakly reflect whether the generated compounds are medicinally plausible and suitable for downstream...

    arxiv.org/abs/2607.13155 · PDF

  65. 65

    The Seriality Gap in Video Diffusion Models

    Jorge Diaz Chao, Konpat Preechakul, Yuxi Liu, Yutong Bai

    cs.LG · cs.CV

    When one ball strikes another, then another, video models should predict the consequences of each bounce. In controlled experiments on multi-ball hard-sphere dynamics, we find that the performance of standard bidirectional video diffusion degrades as the causal chain lengthens, even when provided more denoising steps. In a length-matched single-ball control, where ball-ball interactions are absent, the degradation largely disappears,...

    arxiv.org/abs/2607.13031 · PDF

  66. 66

    TerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale

    Zhouchonghao Wu, Akshay Rangesh, Weixin Li, Wei-Jer Chang, Zachary Lee, Tim Wang, Wei Zhan

    cs.LG · cs.AI · cs.RO

    Training robust autonomous driving agents requires a simulator that is fast enough for reinforcement learning at scale, realistic enough to ground behavior in real-world map structure, and diverse enough to cover the safety-critical long tail that logged data rarely contains. We present TerraZero, a procedural driving simulator and self-play training stack. A configurable C engine runs simulation on the CPU and policy inference on the GPU...

    arxiv.org/abs/2607.13028 · PDF

This edition is part of The Daily Abstract — cs.LG archive. Subscribe to receive these in your inbox each morning, automatically translated to Spanish, with reply-to-PDF: arxivdaily.ignorelist.com.

Colophon Set in Georgia, with system sans for interface chrome and a monospaced stack for code and paper identifiers. Sole accent: amber #D99C5E. Built and served on an always-free VM. The masthead is set 14% letterspaced because newspapers do that and it works.