Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
119 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Confidence Score Based Conformer Speaker Adaptation for Speech Recognition (2206.12045v1)

Published 24 Jun 2022 in eess.AS and cs.SD

Abstract: A key challenge for automatic speech recognition (ASR) systems is to model the speaker level variability. In this paper, compact speaker dependent learning hidden unit contributions (LHUC) are used to facilitate both speaker adaptive training (SAT) and test time unsupervised speaker adaptation for state-of-the-art Conformer based end-to-end ASR systems. The sensitivity during adaptation to supervision error rate is reduced using confidence score based selection of the more "trustworthy" subset of speaker specific data. A confidence estimation module is used to smooth the over-confident Conformer decoder output probabilities before serving as confidence scores. The increased data sparsity due to speaker level data selection is addressed using Bayesian estimation of LHUC parameters. Experiments on the 300-hour Switchboard corpus suggest that the proposed LHUC-SAT Conformer with confidence score based test time unsupervised adaptation outperformed the baseline speaker independent and i-vector adapted Conformer systems by up to 1.0%, 1.0%, and 1.2% absolute (9.0%, 7.9%, and 8.9% relative) word error rate (WER) reductions on the NIST Hub5'00, RT02, and RT03 evaluation sets respectively. Consistent performance improvements were retained after external Transformer and LSTM LLMs were used for rescoring.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (10)
  1. Jiajun Deng (75 papers)
  2. Xurong Xie (38 papers)
  3. Tianzi Wang (37 papers)
  4. Mingyu Cui (31 papers)
  5. Boyang Xue (23 papers)
  6. Zengrui Jin (30 papers)
  7. Mengzhe Geng (42 papers)
  8. Guinan Li (23 papers)
  9. Xunying Liu (92 papers)
  10. Helen Meng (204 papers)
Citations (13)

Summary

We haven't generated a summary for this paper yet.