Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
Twelve mainstream speech generation techniques were used to create fake audios for this dataset. It contains two versions, clean and noisy, with the noisy version created by adding noise from three databases at five different signal-to-noise ratios. The dataset is structured into training, development, and test sets, with a further split into seen and unseen subsets to evaluate model generalization.
Dataset is publicly available with an Open Access (green) license. Baseline source code is hosted on GitHub (https://github.com/ADDchallenge/FAD).