
DEFINE: Exemplar-Guided Accent Control for Zero-Shot TTS
DEFINE decouples speaker identity and accent in zero-shot TTS via separate audio exemplars and a single guidance weight, generalizing accent control beyond training accents with high speaker similarity.
Published Sep 26, 2026 · 0 citations · ▲ 32 on Hugging Face · Code ★ 2
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.





