cs.LG · 2026-08-18 · No. 88

Machine Learning, 2026-08-18.

48 new papers in cs.LG. Titles, authors, abstracts. Links to arXiv. Want this in your inbox every morning? Subscribe →

01 — The papers

48 entries
  1. 01

    Q-based Variational Inverse Reinforcement Learning

    Ondrej Bajgar, Peter Tisnikar, Alessandro Abate, Konstantinos Gatsis, Maike Osborne

    cs.LG

    The development of safe and beneficial AI requires that systems can learn and act in accordance with human preferences. However, explicitly specifying these preferences by hand is often infeasible. Inverse reinforcement learning (IRL) addresses this challenge by inferring preferences, represented as reward functions, from expert behaviour. We introduce Q-based Variational IRL (QVIRL), a novel Bayesian IRL method that recovers a posterior...

    arxiv.org/abs/2608.16888 · PDF

  2. 02

    An Analytical-Prior Framework for Data-Efficient Prediction of Sound-Reduction Frequencies in Rectangular Side-Branch Helmholtz Resonators

    Jiaming Li

    cs.LG

    High-fidelity finite-element simulations can provide accurate numerical predictions for side-branch resonators, but large simulation datasets are expensive to generate and purely data-driven surrogates may become unreliable when simulation-labelled data are scarce. This study develops an analytical-prior learning framework that reuses a low-cost analytical model to improve data efficiency under limited high-fidelity simulation budgets. Two...

    arxiv.org/abs/2608.16873 · PDF

  3. 03

    Data-Efficient and Interpretable Classification of Circulating Tumor Cell Phenotypes in Microfluidic Devices via Deep Learning

    Serena Su, Yifan Wang, Senwei Liang

    cs.LG

    Accurate classification of circulating tumor cell (CTC) phenotypes can provide valuable information for assessing metastatic potential. Label free microfluidic devices provide a hydrodynamic obstacle course that transforms subtle biophysical characteristics of CTCs, including size and deformability, into distinct kinematic trajectories. However, the highly nonlinear fluid structure interactions governing these trajectories make the inverse...

    arxiv.org/abs/2608.16870 · PDF

  4. 04

    Proteus: Incremental Memory Activation for Long-Context Sequence Modeling

    Reza Bayat, Ali Behrouz, Vahab Mirrokni, Aaron Courville

    cs.LG · cs.AI · cs.CL

    The quadratic cost of attention-based sequence models for long contexts has motivated a growing line of research on memory-based models that can compress context into a compact state. However, most existing memory models expose a static memory throughout the entire sequence. Because early tokens face no compression pressure, they occupy too many degrees of freedom and "pollute" the memory state, leaving little capacity for later context and...

    arxiv.org/abs/2608.16844 · PDF

  5. 05

    Time-Aware Validation of Machine Learning Fuel Consumption Models: Evidence from 1\,Hz Operational Data, CCGS \textit{Sir Wilfrid Laurier}

    Samarasimha Reddy Chittamuru, Ayhan Akinturk, Allison Kennedy, Joshua Barnes, Matthew Hamilton

    cs.LG

    Ship fuel consumption (SFC) prediction supports vessel operation optimisation, emissions estimation, and decision support systems (DSS) for sustainable maritime transportation. Numerous data-driven fuel models have been developed over the past two decades, but a critical and often overlooked limitation lies in their validation practices: most studies evaluate performance using random train--test splits, which, applied to high-frequency...

    arxiv.org/abs/2608.16833 · PDF

  6. 06

    CaliBench: Are the Stochastic Dynamics of Video World Models Physically Calibrated?

    Jonathan Sadeghi, Jenny Seidenschwarz, Jesse Allardice, Sirish Srinivasan, Benjamin Graham, Jeffrey Hawke

    cs.LG · cs.AI

    Video world models approximate the stochastic distribution of physical outcomes through generative sampling, but existing benchmarks score individual generations or compare distributions coarsely over a whole dataset, leaving the fine-grained aleatoric uncertainty of specific phenomena untested. We introduce CaliBench, which scores outcomes in a physically interpretable discrete space - a bin index, a die face, a suit, a colour - rather than...

    arxiv.org/abs/2608.16829 · PDF

  7. 07

    GEO-Flag: Detecting and Measuring GEO-Optimized Web Content

    Junjie Chu, Ye Leng, Mingjie Li, Yun Shen, Xinyue Shen, Yang Zhang

    cs.LG · cs.CR · cs.IR

    Generative Engine Optimization (GEO) modifies web content to increase its likelihood of being selected and cited by generative search engines. This can give strategically optimized pages visibility disproportionate to their authority or relevance and even make weak or false information appear well supported. Unlike conventional search, generative search synthesizes information into direct answers rather than presenting competing sources,...

    arxiv.org/abs/2608.16824 · PDF

  8. 08

    Beyond $L_2$: Generalizing Abductive Latent Explanations to Diverse Prototype-Based Architectures

    Jules Soria, Alban Grastien, Romain Xu-Darme, Julien Girard-Satabin, Zakaria Chihani, Daniela Cancila

    cs.LG

    Prototype-based neural networks are hailed as interpretable-by-design architectures. Recently, Abductive Latent Explanations (ALE) were introduced to provide formal, mathematically guaranteed explanations that leverage the intrinsic structure of these networks to ensure both predictive safety and human readability. ALEs rely on computing tight bounds on latent space distances to produce formal explanations. However, existing ALE formulations...

    arxiv.org/abs/2608.16773 · PDF

  9. 09

    On the Principles Behind Neural Network Optimizers

    Yushun Zhang

    cs.LG · math.OC

    Reliable optimization is central to neural network (NN) training, yet Adam, the default optimizer for modern LLMs, rests on a fragile foundation. This thesis develops a principled grounding for Adam and motivates new designs. First, we revisit Adam's divergence--convergence debate and show the existence of a problem-dependent phase transition: with properly chosen, batch-size-dependent hyperparameters, Adam converges, whereas under...

    arxiv.org/abs/2608.16760 · PDF

  10. 10

    Would this change your answer? Evaluating Explanations of LLM Behavior In The Wild with Counterfactual Experiments

    Adam Karvonen, Euan Ong, Subhash Kantamneni, Samuel Marks

    cs.LG · cs.AI

    Many areas of AI research, such as language model interpretability and chain of thought faithfulness, seek to explain model behaviors. But what constitutes a "good" explanation? In this work, we evaluate explanations through the lens of counterfactual simulatability-whether the explanation is useful for predicting model behaviors on related counterfactual inputs. To this end, we introduce CHIVE (Counterfactual Hypothesis Investigation Via...

    arxiv.org/abs/2608.16747 · PDF

  11. 11

    Le Critique: Privileged Value Functions for LLM Reinforcement Learning

    Siddarth Venkatraman, Matthieu Dinot, Laurence Aitchison

    cs.LG

    Reinforcement learning algorithms for Large Language Models (LLMs) are largely distinguished by their variance reduction strategy. Group-relative methods like GRPO reduce gradient variance by sampling multiple rollouts per prompt, but provide only sequence-level credit. Training is also blocked by straggler rollouts, reducing throughput and increasing off-policyness. Learned value functions theoretically address both problems, providing...

    arxiv.org/abs/2608.16739 · PDF

  12. 12

    The Ethical Decision Head: Operationalizing Normative Ethics in Autonomous Vehicles via Reinforcement Learning from Human Feedback

    Thomas Mbrice, Ammar Ali, Sami Mian, Khai Hern Low, Eric Chen, Arshia Aghajani, Wolf Schäfer, Amin Shirangi

    cs.LG

    As autonomous vehicles (AVs) approach Level 4 and Level 5 operational capability [SAE International, 2018], their on- board decision systems must handle not only safety-critical locomotion but also their subsequent moral weight. This paper details the Ethical Decision Head (EDH), a deep re- inforcement learning (RL) framework that encodes ethical reasoning as a differentiable reward signal, enabling a pol- icy gradient agent to learn...

    arxiv.org/abs/2608.16710 · PDF

  13. 13

    Learning to Unlearn: Machine Unlearning via Learning the Unlearning Behaviors

    Hang Zhang, Kaifeng Zhang, Yixiao Ma, Weijie Xu, Ye Zhu, Kai Ming Ting

    cs.LG · cs.AI

    Various machine unlearning techniques have been developed in response to privacy legislation requirements, enabling individuals to exercise their legal right to have their data $D_f$ removed from a machine learning model. This process is typically accomplished via the use of an unlearning function denoted as $U$. Existing methods focus on designing an intricate $U$ to unlearn $D_f \subset D$ from a previous model $A(D)$, so that the unlearned...

    arxiv.org/abs/2608.16700 · PDF

  14. 14

    UniTAC: Universal Task-Aware Compression via Weighted Distortion Measures

    Homa Esfahanizadeh, Matin Mortaheb, Jinfeng Du, Harish Viswanathan

    cs.LG · cs.AI · cs.IT · cs.MM

    Physical AI systems such as autonomous vehicles and robots rely on timely exchange of high-dimensional sensory signals under tight bandwidth, latency, and energy budgets. Because the task driving downstream decisions evolves over time, a task-specific codec is brittle and retraining one per task is infeasible in the field. We propose UniTAC, a single learned image codec spanning universal (task-agnostic) to task-specialized operation,...

    arxiv.org/abs/2608.16696 · PDF

  15. 15

    Hoeffding adaptive splitting trees for data stream classification with concept drift and ensemble learning

    Daniel Nowak Assis, Jean Paul Barddal, Fabrício Enembreck

    cs.LG · cs.AI

    Ensembles of decision trees are well-established methods for data stream classification. In ensemble learning, Hoeffding Trees are widely adopted as base learners, performing periodic split attempts according to the Hoeffding bound. Recent studies, however, indicate that this standard splitting mechanism lacks adaptability, while adaptive trees that trigger splits in response to performance degradation have achieved superior results. In this...

    arxiv.org/abs/2608.16659 · PDF

  16. 16

    Variational Outlier-Robust Gaussian Process Regression with Generative Modeling

    Arslan Majal, Aamir Hussain Chughtai

    cs.LG

    Outliers can substantially distort Gaussian process regression (GPR) due to its conventional Gaussian observation likelihood, leading to inaccurate model learning and prediction. To address this limitation, this article introduces a generative GPR model that captures observation-specific contamination and adaptively mitigates the influence of outliers. Subsequently, a variational generalized expectation-maximization procedure is used to learn...

    arxiv.org/abs/2608.16606 · PDF

  17. 17

    Learning Generalizable Reconstruction of High-Dimensional Neural Dynamics

    Anima Kujur, Zahra Monfared

    cs.LG

    Accurate reconstruction of long-duration neural recordings is challenging because local field potentials (LFPs) are high-resolution, multichannel, transient, and variable across subjects. We present PCA-DMD, a scalable operator-theoretic framework that segments LFP recordings into overlapping windows, projects them into a compact PCA space, learns linear Koopman evolution in the latent space, and reconstructs continuous signals through...

    arxiv.org/abs/2608.16569 · PDF

  18. 18

    One Residual with Three Reuses: A Wristband Front End for Gesture Sensing

    Sam Rifaki

    cs.LG · cs.AR · cs.HC

    Continuous wrist-worn hand sensing for gesture interfaces and motor symptom monitoring needs an always-on front end that fits inside a coin-cell power budget while pairing a micro-electro-mechanical-systems (MEMS) inertial measurement unit (IMU) with a 60 GHz frequency-modulated continuous-wave (FMCW) radar to stay robust under occlusion and on-body drift. We present a design study of such a wristband front end in which classifier wake-up...

    arxiv.org/abs/2608.16542 · PDF

  19. 19

    When Tool-Backed Skill Retrieval Fails: Source-Style Collapse in Executable Capability Retrieval

    Yiqi Liu, Joseph James, Yang Wang, Chenghao Xiao, Chenghua Lin

    cs.LG · cs.IR

    Large-scale agents increasingly rely on retrieval to access external capabilities. We study this retrieval gate in structured tools and APIs, a measurable class of tool-backed executable skills that must be surfaced before an agent can plan, incorporate, or act. In this setting the retrieval layer can silently fail even when the capability corpus is fixed: on ToolRet, a retriever fine-tuned on one source-specific slice collapses on another...

    arxiv.org/abs/2608.16502 · PDF

  20. 20

    Graph Machine Learning: An Opportunity for Power Systems

    Martin Sadric, Sebastian Pütz, Christian Nauck, Veit Hagenmeyer, Frank Hellmann, Dirk Witthaut, Benjamin Schäfer

    cs.LG · cs.AI · cs.CE · eess.SY

    Modern power systems face growing operational complexity driven by the integration of renewable energy sources, decentralization, and the need for real-time decision-making across a wide range of timescales. Addressing these challenges traditionally relies on model-based methods that, while accurate, can be too slow for operational demands. Machine learning (ML) has therefore emerged as a faster, data-driven alternative. As grid topology...

    arxiv.org/abs/2608.16494 · PDF

  21. 21

    Pallas: A Proactive KV Cache Migration Framework for LLM Inference in AI-RAN

    Tianhang Ding, Jianchun Liu, Hongli Xu

    cs.LG

    AI-RAN brings large language model (LLM) serving close to mobile users, but cellular handover can separate an active request from its inference state: the user attaches to a target base station (gNB) while the large and growing key-value (KV) cache remains at the source. Retaining inference at the source preserves service continuity but persistently increases inter-token latency (ITL), whereas recovering the state at the target restores...

    arxiv.org/abs/2608.16477 · PDF

  22. 22

    Reference-free logged energy-oracle recovery for neural approximations of symmetric coercive variational problems: conforming Riesz reconstruction and archive-level selection

    Karim Bounja, Lahcen Laayouni, Boujemaa Achchab, Abdeljalil Sakat

    cs.LG · math.NA

    Neural PDE training yields a finite checkpoint archive, yet its logged energy errors are inaccessible without the exact solution, while loss-based selection does not necessarily recover the logged energy oracle. For admissible neural approximations of symmetric coercive variational problems, we introduce a reference-free selection rule based on minimizing a computable conforming Riesz monitor. The exact residual-energy identity and conforming...

    arxiv.org/abs/2608.16473 · PDF

  23. 23

    Localized TabICLv2: Scaling Tabular In-Context Learning through k-NN

    Beimnet Bekele Guta

    cs.LG

    Foundational models for tabular data have made significant progress in recent years, with TabICLv2 reporting state-of-the-art performance on several tabular classification tasks. However, full-context tabular ICL still suffers from attention cost that grows with the training-context size, which limits its ability to handle large datasets efficiently. Localized TabICLv2 introduces a method that reduces the inference cost of TabICLv2 by...

    arxiv.org/abs/2608.16429 · PDF

  24. 24

    PertMind: Eliciting Emergent Biological Reasoning in LLM via Reinforcement Learning on Cellular Perturbation Data

    Zhenchao Tang, Xiaogang Xu, Tianxu Lv, Jiahui Guan, Jiale Zhou, Haohuai He, Zhi Song, Hanbo Huang, Jiehui Huang,...

    cs.LG · cs.AI · q-bio.QM

    Large language models can describe mechanisms, yet scalable post-training still depends on costly, manually curated biological reasoning traces. Here we show that cellular perturbation atlases can instead become reinforcement-learning environments, where measured gene responses provide computable rewards for biological reasoning. We introduce PertMind, which combines trusted-trajectory supervised initialization with gene-, pathway-, and...

    arxiv.org/abs/2608.16419 · PDF

  25. 25

    Evolving Executable Pipeline Programs for AutoML with Language Models

    Sofoklis Kitharidis, Cor J. Veenman, Jan N. van Rijn, Thomas Bäck, Niki van Stein

    cs.LG · cs.NE

    Automated machine learning (AutoML) systems search for pipelines within a space of preprocessing operators, learners, and hyper-parameters specified in advance: they can select and tune known components, but cannot produce structure outside that space. We present LACE, an AutoML framework that instead searches over complete executable pipeline programs: an evolutionary loop maintains a population of scikit-learn-compatible Python classes, and...

    arxiv.org/abs/2608.16416 · PDF

  26. 26

    TRACE-CASH: Trial-History-Conditioned Reinforcement Learning for Adaptive Configuration Exploration in Time-Series CASH

    Yu-Han Huang, Yujia Wu, Vincent S. Tseng

    cs.LG

    Combined algorithm selection and hyperparameter optimization (CASH) searches a conditional space in which the selected model determines which hyperparameters are active. In time-series forecasting, temporal choices, chronological validation, and costly evaluations further complicate this search. Controlled comparisons of heterogeneous search methods under a shared time-series CASH (TS-CASH) evaluation protocol remain limited. Within this...

    arxiv.org/abs/2608.16410 · PDF

  27. 27

    SoftModel: A Neural Model That Grows Its Own Topology -- Governed Structural Growth for Continual In-Service Learning

    Zhoumin Xie

    cs.LG

    Today, a neural system is almost always used in two phases -- trained, then deployed -- and in that regime it freezes twice: training ends, and the topology itself was never a degree of freedom. We take the opposite premise as an axiom -- total plasticity: no part of a model, including its structure, is ever frozen -- and derive the governance a lifelong learner then requires. The design's target regime is continual, in-service learning: a...

    arxiv.org/abs/2608.16409 · PDF

  28. 28

    FETERS: Few-Shot Early Time-Series Classification via Effective Ratio Selection

    Chen-An Tai, Yujia Wu, Vincent S. Tseng

    cs.LG

    Early time-series classification (ETSC) aims to make accurate predictions from partially observed time series as early as possible. Although various stopping mechanisms and feature learning strategies have been developed for ETSC, most existing methods assume access to sufficient labeled training data, which may be unrealistic in applications with limited annotation. Under limited supervision, learning an additional sample-level stopping...

    arxiv.org/abs/2608.16385 · PDF

  29. 29

    Coverage-Maximizing Multinomial Subset Routing under Operational Constraints

    Quan Zhou, Yiyan Huang

    cs.LG · cs.AI

    We introduce Multinomial Subset Routing (MSR), a new online routing framework over $K$ experts in which the learner keeps a multinomial routing policy instead of a deterministic subset of experts. At each round, the learner samples $M$ experts i.i.d. from the multinomial policy, and the resulting set of distinct sampled experts forms the routed subset. The reward depends only on the best-performing expert(s) in the routed subset. This reward...

    arxiv.org/abs/2608.16375 · PDF

  30. 30

    OceanDepths: A Global Dataset of Paired Subsurface and Surface Ocean Observations

    Simon Donike, Ruben Cartuyvels, Antonino Ian Ferola, Elisa Carli, Diego Fernandez Prieto, Marie-Helene Rio

    cs.LG · cs.AI · cs.CV

    Despite comprising over 70\% of its surface, the world's oceans are critically underobserved compared to the land surface or the atmosphere.Understanding the global ocean requires jointly observing its surface and subsurface structure, yet no standardized, high-resolution dataset couples satellite surface fields to co-located \emph{in situ} depth profiles in an AI-ready format.Existing resources either consist of model-reconstructed gridded...

    arxiv.org/abs/2608.16373 · PDF

  31. 31

    Task-Anchored Representation Shaping for Pre-Trained Model-Based Continual Learning

    Zhiming Xu, Huiyu Yi, Zhen-Hao Xie, Baile Xu, Furao Shen, Jian Zhao, Suorong Yang

    cs.LG

    Pre-trained models (PTMs) provide a strong foundation for continual learning by offering stable representations that facilitate lightweight adaptation to new tasks. However, adapting well to each task does not ensure reliable inference over all learned tasks. Since task boundaries are often artificial and semantically entangled, an input from an unknown task can remain ambiguous even with strong PTM features, making cross-task prediction a...

    arxiv.org/abs/2608.16345 · PDF

  32. 32

    Transfer Learning of Keystroke Dynamics for Cross-Device User Authentication

    Nuwan Kaluarachchi, Sevvandi Kandanaarachchi, Kristen Moore, Arathi Arakala, Conrad Sanderson

    cs.LG · cs.HC

    Keystroke dynamics (typing patterns) can be used as a behavioural biometric modality for user authentication, with applications such as fraud prevention. While the modality has been shown to work well for single device authentication, its application to cross-device scenarios is more challenging. Dynamics learned on one device (eg., phone) may not be directly applicable to authentication on a secondary device with a different form factor...

    arxiv.org/abs/2608.16334 · PDF

  33. 33

    Advancing Open and Reproducible Relational Learning: RelArena-$α$, TabPFN-Rel and RPI

    Adrian Hayler, Klemens Flöge, Alan Arazi, Rishabh Ranjan, Jure Leskovec, Felix Birkel, Brendan Roof, Anurag Garg,...

    cs.LG

    This first release of Prior Labs in relational learning shows our continued commitment to open science. We open-source three pieces of software that we expect to accelerate research in the field towards meaningful real-world impact. We aim to steer further development based on feedback from, and in collaboration with, the community. Given the early stage of development, our $α$-release targets researchers and early-adopting practitioners....

    arxiv.org/abs/2608.16319 · PDF

  34. 34

    SCALE: State-Calibrated Latent Embeddings for JEPA Planning in the Right Geometry

    Jiaming Hu, Yan Zheng, Tian Wang

    cs.LG

    Joint-embedding predictive world models plan by scoring predicted terminal embeddings against a goal embedding using a cost defined on the representation itself. Two prominent strategies for obtaining non-collapsed representations are to inherit a pretrained feature space, as in DINO-WM, and to learn an embedding end to end with anti-collapse regularization, as in LeWorldModel (LeWM) with SIGReg. These strategies show complementary strengths...

    arxiv.org/abs/2608.16287 · PDF

  35. 35

    Foresight-England: Development of a National-Scale Generative AI Model of Electronic Health Records for Medical Event Prediction across the COVID-19 Pandemic

    Simon Ellershaw, Christopher Tomlinson, Zeljko Kraljevic, Spiros Denaxas, Harry Hemingway, Cathie Sudlow, Angela M....

    cs.LG · cs.AI

    Foresight-England (Foresight-E) is the first national-scale generative foundation model of electronic health records (EHRs), developed as a research pilot strictly for COVID-19 research. We evaluated its ability to model the direct and indirect effects of the pandemic. Trained from scratch entirely within the NHS England Secure Data Environment, Foresight-E is a 243-million-parameter transformer decoder. It was trained and evaluated on...

    arxiv.org/abs/2608.16273 · PDF

  36. 36

    Efficient Coreset Selection via K-Nearest Neighbor Graphs

    Yingfan Liu, Leiyu Zhang, Jiadong Xie, Mingzhe Wang, Jeffrey Xu Yu, Jiangtao Cui

    cs.LG

    Coreset selection reduces the cost of model training by replacing a large training set with a small representative subset. Existing gradient-approximation coreset methods such as CRAIG and cluster-based variants can preserve model accuracy. Still, their selection stages often rely on dense pairwise distances or large item-cluster bound matrices, leading to high time and memory costs on large datasets. This paper proposes KNNG-CS, a...

    arxiv.org/abs/2608.16270 · PDF

  37. 37

    SAUL: Sharpness-Aware Augmented-Lagrangian Unlearning

    Jaewan Choi, Junyoung Yang, Sangdon Park

    cs.LG

    Machine unlearning in Large Language Models (LLMs) faces a critical trade-off between erasing target knowledge and preserving general utility. We propose SAUL (Sharpness-Aware Augmented-Lagrangian Unlearning), which formulates unlearning as a constrained minimization problem following the principle of "forget enough, but no more than necessary." At its core, SAUL formulates forgetting as an explicit constraint with a prescribed satisfaction...

    arxiv.org/abs/2608.16249 · PDF

  38. 38

    The Trade-off Between Covariate Dependence and Latent Structure in Representation Learning

    Małgorzata Łazęcka, Ewa Szczurek

    cs.LG

    Disentangled representation learning seeks latent representations whose indicidual dimensions each align with a distinct covariate. Unsupervised approaches typically target latent dimension independence, yet this gives no guarantee that the resulting dimensions align with semantically meaningful covariates. Supervised approaches structure the latent space using observed covariates, but under correlated covariates they cannot simultaneously...

    arxiv.org/abs/2608.16245 · PDF

  39. 39

    Optimizing Multi-Market Participation of Battery and Electrolyser Systems Based on Field Performance

    Chunyang Zhao, Stoyan Trenchev, Shi You, Chresten Træholt

    cs.LG

    The increasing share of renewable energy in power systems creates a need for fast-response and flexible resources to maintain system stability. With the expansion of electricity markets and ancillary service products, opportunities arise to stack revenues across multiple services. Long-term Power-to-X (PTX) electrolysers and short-term battery energy storage systems (BESS) are prevalent flexible resources, yet most studies neglect real...

    arxiv.org/abs/2608.16238 · PDF

  40. 40

    A Privacy Study of Sparse Collaborative Inference

    Maximilian Andreas Hoefler, Karsten Mueller, Wojciech Samek

    cs.LG

    Collaborative inference (CI) splits a model between an edge device and a server, whereby the client computes an intermediate activation, transmits it, and the server completes the computation. This raises two concerns, the communication cost of the transmission and the risk that it reveals private information about the input. Recent work reduces this cost by sparsifying activations and entropy-coding the result. Sparsity has also been argued...

    arxiv.org/abs/2608.16236 · PDF

  41. 41

    Beyond Peak Backlog: Conditional Energy and Temporal Geometry in Capacity-Constrained Delayed Bandit Optimization

    Anling Xiang, Yuwen Yang, Yang Shen

    cs.LG

    What is the right delay complexity when a learner can track only $C$ pending feedback items and discarded feedback is permanently lost? Existing one-point bandit convex optimization guarantees in this model pay $\sqrt{Tσ_{\max}}$, where $σ_{\max}$ is the peak backlog, although unlimited tracking admits the sharper $\sqrt{d_{\mathrm{tot}}}$ dependence on total delay. We introduce a scheduler-side conditional-energy interface that separates...

    arxiv.org/abs/2608.16216 · PDF

  42. 42

    Quantifying the Gap Between Laboratory Battery Test Patterns and Field Duty Profiles

    Chunyang Zhao, Chresten Træholt

    cs.LG

    Laboratory battery tests provide the main empirical basis for battery performance and degradation studies, but their operating patterns do not directly represent field duty profiles. This paper quantifies the gap by comparing six accessible evidence sources covering controlled cycling, drive-cycle testing, dynamic cycling, NMC811 laboratory ageing, a real electric-vehicle charging trace, and fleet-scale electric-vehicle state-of-health (SOH)...

    arxiv.org/abs/2608.16212 · PDF

  43. 43

    Conditional Evaluation of Language Models with Cheap Auxiliary Signals

    Zhi Zhang, Lingfeng Lyu, Yue Kang, Doudou Zhou

    cs.LG · stat.ML

    Aggregate accuracy hides where models succeed and fail. Estimating conditional performance profiles from gold labels alone is expensive, while cheap auxiliary signals such as LLM-judge scores, pairwise comparisons, confidence scores, and judge-disagreement features can be collected for every benchmark item but are often biased or miscalibrated. We propose LACE (Local Augmented Control-Variate Evaluation), a semi-supervised estimator for...

    arxiv.org/abs/2608.16210 · PDF

  44. 44

    Multi-Granularity Sentiment Integration for LLM-Based Multimodal Sentiment Analysis

    Shanshan Lin, Yuesheng Wu, Chao Chen, Yizhe Yang, Zhihao Chen, Zexian Yang, Xiangwen Liao

    cs.LG

    Multimodal sentiment analysis (MSA) aims to predict sentiment polarity and intensity from heterogeneous inputs such as text, audio, and vision. While large language models (LLMs) offer strong semantic priors for MSA, effectively incorporating audio and visual signals effectively remains challenging. A key challenge is that audio and visual sentiment cues evolve over different temporal scales, yet many LLM-based methods compress these signals...

    arxiv.org/abs/2608.16201 · PDF

  45. 45

    Understanding and Stabilizing Deep Q-Learning via Controlled Bootstrapping and Regulated Value Dynamics

    Bozhou Chen, Yongyi Wang, Hanyu Liu, Xionghui Yang, Wenxin Li

    cs.LG · cs.AI

    Deep Q-learning (DQL) has achieved remarkable empirical success in reinforcement learning, yet its training process remains notoriously unstable. Existing studies often attribute instability to isolated factors such as overestimation bias or representation learning issues, lacking a unified understanding of how different sources of instability interact during recursive value estimation. In this work, we provide a systematic analysis of...

    arxiv.org/abs/2608.16182 · PDF

  46. 46

    Demystifying Oversmoothing in Sheaf Neural Networks: An Index-Theoretic Criterion

    Junwen Dong, Yuhan Peng, Hao Li, Huitao Feng, Kelin Xia

    cs.LG · math.DG

    To combat oversmoothing in Graph Convolutional Networks, Sheaf Neural Networks (SNNs) were proposed as a generalization by equipping the graph with a sheaf structure and replacing the graph Laplacian with a sheaf Laplacian $\mathcal{L}$. Existing analyses connect sheaf diffusion to oversmoothing via the harmonic space ($\ker\mathcal{L}$), taking its absolute dimension as an indicator of anti-oversmoothing capacity. However, absolute dimension...

    arxiv.org/abs/2608.16180 · PDF

  47. 47

    A Tree-Structured Approach for Phishing Template and Attacker Attribution Analysis

    Unai Agirre, Imanol Jerico, Felipe Castaño, Andrea Venturi, Francesco Zola

    cs.LG · cs.AI · cs.ET

    Phishing remains a persistent and evolving cybersecurity threat, with attack volumes reaching record levels. This growth is driven by the industrialization of phishing through widely available phishing kits and reusable templates, which enable cybercriminals to rapidly generate and deploy large numbers of fraudulent webpages. Although surface-level attributes may differ across these websites, their underlying structures often exhibit...

    arxiv.org/abs/2608.16158 · PDF

  48. 48

    REFLEX: Reflexive Equilibrium Fixed-point Learning for Endogenous eXchanges

    Vignesh Nagarajan, Shriraghav Ashok

    cs.LG · cs.CE · cs.GT

    In over-the-counter corporate bond markets, dealers compete for client trades by quoting bid and ask prices. Tighter quotes attract more business, but also informed customers more likely to trade ahead of adverse price moves, leaving the dealer holding the risk. As dealers increasingly use machine learning to set quotes, they retrain these models on the trades their own quotes attract, creating a feedback loop in which each model reshapes the...

    arxiv.org/abs/2608.16155 · PDF

This edition is part of The Daily Abstract — cs.LG archive. Subscribe to receive these in your inbox each morning, automatically translated to Spanish, with reply-to-PDF: arxivdaily.ignorelist.com.

Colophon Set in Georgia, with system sans for interface chrome and a monospaced stack for code and paper identifiers. Sole accent: amber #D99C5E. Built and served on an always-free VM. The masthead is set 14% letterspaced because newspapers do that and it works.