Sign in to view source links and access this dataset
Description
IndicTTS-p1 is a dataset for text-to-speech synthesis, published on Kaggle. The title suggests it contains data for Indic languages, which likely includes audio recordings and corresponding text transcripts. The dataset's specific size, languages, and collection details are not provided in the available metadata.
Use Cases
Training a text-to-speech model for an Indic language (inferred from domain, verify after download)
Benchmarking speech synthesis quality across different languages (inferred from domain, verify after download)
Creating a voice cloning or voice conversion system for low-resource languages (inferred from domain, verify after download)
Strengths
Published on Kaggle, a platform with established data sharing and versioning tools.
Limitations
Metadata is minimal; actual content requires verification after download.
Column-level documentation is absent; field semantics must be inferred after download.
Row count, file formats, and license are unknown, which may limit suitability assessment.
License is unknown; users must verify permissible usage after download.