Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
97 tokens/sec
GPT-4o
53 tokens/sec
Gemini 2.5 Pro Pro
44 tokens/sec
o3 Pro
5 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Senone-aware Adversarial Multi-task Training for Unsupervised Child to Adult Speech Adaptation (2102.11488v1)

Published 23 Feb 2021 in cs.SD, cs.AI, and eess.AS

Abstract: Acoustic modeling for child speech is challenging due to the high acoustic variability caused by physiological differences in the vocal tract. The dearth of publicly available datasets makes the task more challenging. In this work, we propose a feature adaptation approach by exploiting adversarial multi-task training to minimize acoustic mismatch at the senone (tied triphone states) level between adult and child speech and leverage large amounts of transcribed adult speech. We validate the proposed method on three tasks: child speech recognition, child pronunciation assessment, and child fluency score prediction. Empirical results indicate that our proposed approach consistently outperforms competitive baselines, achieving 7.7% relative error reduction on speech recognition and up to 25.2% relative gains on the evaluation tasks.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (2)
  1. Richeng Duan (1 paper)
  2. Nancy F. Chen (97 papers)
Citations (7)

Summary

We haven't generated a summary for this paper yet.