No AI summary available for this article.
Why It Matters
Voice actors often re-read the same script while modifying their delivery in response to performance directions.
Provenance
Discovered via ArXiv and published by ArXiv.
Key Claims
Original description
Voice actors often re-read the same script while modifying their delivery in response to performance directions. We study this setting as direction-following TTS, where a system generates a new utterance that reflects a given direction relative to a reference utterance while preserving speaker identity and linguistic content. A key challenge is the lack of training data capturing such relative modifications. To address this, we propose a scalable pseudo-triplet construction pipeline that generates~(reference utterance, direction text, modified utterance) triplets. It generates controlled style...
Discovered via ArXiv
Research papers and preprints from arXiv.
Publisher: arxiv.org
ID: http://arxiv.org/abs/2609.02623v1 · Indexed about 1 hour ago