cs.SD · 2026-07-20 · No. 59

Sound, 2026-07-20.

2 new papers in cs.SD. Titles, authors, abstracts. Links to arXiv. Want this in your inbox every morning? Subscribe →

01 — The papers

2 entries
  1. 01

    AuEmoChat: Authentic Emotion Understanding and Rendering for Conversational Speech Synthesis

    Zhenqi Jia, Yuan Zhao, Aruukhan, Rui Liu, Haizhou Li

    cs.SD · cs.AI

    Conversational Speech Synthesis (CSS) aims to synthesize speech with human-like emotional expression and contextual consistency in user-agent interactions. Existing CSS methods struggle to render authentic human emotions due to limited predefined emotion label spaces (e.g., seven emotion categories), while redundant multimodal tokens in multi-turn dialogue history interfere with context understanding. To address these issues, we propose...

    arxiv.org/abs/2607.15755 · PDF

  2. 02

    SpeechGuard: Online Defense against Backdoor Attacks on Speech Recognition Models

    Jinwen Xin, Xixiang Lv

    cs.SD · cs.CR · cs.LG

    Backdoor attacks pose a critical threat to neural network models, allowing attackers to implant a backdoor during the training phase by manipulating a small portion of the training data. In security-sensitive applications such as voice interaction for autonomous driving, the presence of backdoor attacks introduces substantial security risks. This study focuses on implementing backdoor defense measures for speech recognition models in...

    arxiv.org/abs/2607.15697 · PDF

This edition is part of The Daily Abstract — cs.SD archive. Subscribe to receive these in your inbox each morning, automatically translated to Spanish, with reply-to-PDF: arxivdaily.ignorelist.com.

Colophon Set in Georgia, with system sans for interface chrome and a monospaced stack for code and paper identifiers. Sole accent: amber #D99C5E. Built and served on an always-free VM. The masthead is set 14% letterspaced because newspapers do that and it works.