Papers

Topics

Authors

Recent

View all

Detailed Answer

Quick Answer

Concise responses based on abstracts only

Detailed Answer

Well-researched responses based on abstracts and relevant paper content.

Custom Instructions Pro

Preferences or requirements that you'd like Emergent Mind to consider when generating responses

Gemini 2.5 Flash

Gemini 2.5 Flash 88 tok/s

Gemini 2.5 Pro 52 tok/s Pro

GPT-5 Medium 12 tok/s Pro

GPT-5 High 19 tok/s Pro

GPT-4o 110 tok/s Pro

GPT OSS 120B 470 tok/s Pro

Kimi K2 197 tok/s Pro

2000 character limit reached

MPO: Multidimensional Preference Optimization for Language Model-based Text-to-Speech (2509.00685v1)

Published 31 Aug 2025 in eess.AS and cs.SD

Abstract: In recent years, text-to-speech (TTS) has seen impressive advancements through large-scale LLMs, achieving human-level speech quality. Integrating human feedback has proven effective for enhancing robustness in these systems. However, current approaches face challenges in optimizing TTS with preference data across multiple dimensions and often suffer from performance degradation due to overconfidence in rewards. We propose Multidimensional Preference Optimization (MPO) to better align TTS systems with human preferences. MPO introduces a preference set that streamlines the construction of data for multidimensional preference optimization, enabling alignment with multiple dimensions. Additionally, we incorporate regularization during training to address the typical degradation issues in DPO-based approaches. Our experiments demonstrate MPO's effectiveness, showing significant improvements in intelligibility, speaker similarity, and prosody compared to baseline systems.

Collections

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Paper Prompts

Explore 10 Community Prompts

Follow-up Questions

We haven't generated follow-up questions for this paper yet.

Generate Now

MPO: Multidimensional Preference Optimization for Language Model-based Text-to-Speech (2509.00685v1)

Collections

Summary

Paper Prompts

Follow-up Questions

Authors (4)

Don't miss out on important new AI/ML research

MPO: Multidimensional Preference Optimization for Language Model-based Text-to-Speech (2509.00685v1)

Collections

Summary

Paper Prompts

Follow-up Questions

Related Papers

Authors (4)

Don't miss out on important new AI/ML research