Papers

Topics

Authors

Recent

View all

Assistant

AI Research Assistant

Well-researched responses based on relevant abstracts and paper content.

Custom Instructions Pro

Preferences or requirements that you'd like Emergent Mind to consider when generating responses.

Gemini 2.5 Flash

Gemini 2.5 Flash 73 tok/s

Gemini 2.5 Pro 41 tok/s Pro

GPT-5 Medium 32 tok/s Pro

GPT-5 High 35 tok/s Pro

GPT-4o 84 tok/s Pro

Kimi K2 185 tok/s Pro

GPT OSS 120B 441 tok/s Pro

Claude Sonnet 4.5 36 tok/s Pro

2000 character limit reached

A Language Modeling Approach to Diacritic-Free Hebrew TTS (2407.12206v1)

Published 16 Jul 2024 in cs.CL, cs.SD, and eess.AS

Abstract: We tackle the task of text-to-speech (TTS) in Hebrew. Traditional Hebrew contains Diacritics, which dictate the way individuals should pronounce given words, however, modern Hebrew rarely uses them. The lack of diacritics in modern Hebrew results in readers expected to conclude the correct pronunciation and understand which phonemes to use based on the context. This imposes a fundamental challenge on TTS systems to accurately map between text-to-speech. In this work, we propose to adopt a LLMing Diacritics-Free approach, for the task of Hebrew TTS. The model operates on discrete speech representations and is conditioned on a word-piece tokenizer. We optimize the proposed method using in-the-wild weakly supervised data and compare it to several diacritic-based TTS systems. Results suggest the proposed method is superior to the evaluated baselines considering both content preservation and naturalness of the generated speech. Samples can be found under the following link: pages.cs.huji.ac.il/adiyoss-lab/HebTTS/

Citations (1)

View on Semantic Scholar