Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
WordVoice-5A is a large-scale bilingual dataset containing approximately 4.7k hours of Mandarin and English speech with fine-grained word-level acoustic annotations. It is designed for high-precision controllable Text-to-Speech research and was created by author XXH333. The dataset page was last updated on 2026-06-27.
License is unknown; terms of use must be verified before application.