eess.AS · 2026-08-23 · No. 93

Audio and Speech Processing, 2026-08-23.

2 new papers in eess.AS. Titles, authors, abstracts. Links to arXiv. Want this in your inbox every morning? Subscribe →

01 — The papers

2 entries
  1. 01

    $TCP_α$: Margin-Controlled Confidence estimation for reliable Music Information Retrieval

    Parampreet Singh, Anushka Singh, Sumit Kumar, Vipul Arora

    eess.AS · cs.LG

    Deep neural networks are often overconfident, assigning high confidence even to incorrect predictions. Consequently, users lack a reliable signal for deciding when a prediction can be trusted. Post-hoc confidence estimation addresses this by training a lightweight auxiliary head over a frozen classifier. Existing targets, however, suffer from inherent ambiguity: they assign overlapping confidence values to correct and incorrect predictions,...

    arxiv.org/abs/2608.20326 · PDF

  2. 02

    Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio Learners

    Umberto Cappellazzo, Xubo Liu, Stavros Petridis, Maja Pantic

    eess.AS · cs.AI · cs.SD

    Self-supervised learning (SSL) has driven substantial progress in audio representation learning, though existing methods have increasingly relied on elaborate pre-training recipes to reach competitive performance. A markedly different pre-training philosophy underpins the most influential progress in language modeling and, more recently, in visual representation learning: rather than train encoders as static feature extractors, models are...

    arxiv.org/abs/2608.19863 · PDF

This edition is part of The Daily Abstract — eess.AS archive. Subscribe to receive these in your inbox each morning, automatically translated to Spanish, with reply-to-PDF: arxivdaily.ignorelist.com.

Colophon Set in Georgia, with system sans for interface chrome and a monospaced stack for code and paper identifiers. Sole accent: amber #D99C5E. Built and served on an always-free VM. The masthead is set 14% letterspaced because newspapers do that and it works.