Existence of per-word emotion-annotated speech data
Determine whether annotated training data currently exists in which emotions are labeled separately for each word, rather than for entire sentences or passages, to support per-word emotion tagging in Deaf-centric text-to-speech systems.
References
Per-word, as opposed to per-sentence, emotion tagging requires annotated training data that has emotions annotated separately for each word, rather than entire sentences or passages. It is unclear whether such data currently exists.
— Seeing the Voice, Preserving the Self: A Participatory Design Approach to Deaf-Centric Text-to-Speech
(2609.10199 - Atemnkeng et al., 9 Sep 2026) in Section 5, “Technical Requirements for Implementing the Designs,” subsection “Emotion Customization”