Thorsten-Voice Dataset 2021.02 (Neutral)
- Number of recordings
- 22,668
- Audio duration
- 23+ hours
- Sample rate
- 22,050 Hz
- Channels
- Mono
- Normalization
- -24 dB
- Sentence length (min/avg/max)
- 2 / 52 / 180 characters
- Speaking rate (avg)
- 14 characters/second
- Questions
- 2,780
- Exclamations
- 1,840
Citation (BibTeX)
@dataset{muller_thorsten_2021_5525342,
author = {Müller, Thorsten and Kreutz, Dominik},
title = {Thorsten - Open German Voice (Neutral) Dataset},
month = feb,
year = 2021,
publisher = {Zenodo},
version = {3.0},
doi = {10.5281/zenodo.5525342},
url = {https://doi.org/10.5281/zenodo.5525342}
} Thorsten-Voice Dataset 2021.06 (Emotional)
300 different sentences, each spoken in eight emotions: neutral, disgusted, angry, amused, surprised, sleepy, whispering and drunk (acted only – I was sober during the recordings).
- Number of recordings
- 2,400
- Channels
- Mono
- Normalization
- -24 dB
- Sentence length (min/max)
- 59 / 148 characters
Citation (BibTeX)
@dataset{muller_thorsten_2021_5525023,
author = {Müller, Thorsten and Kreutz, Dominik},
title = {Thorsten - Open German Voice (Emotional) Dataset},
month = jun,
year = 2021,
publisher = {Zenodo},
version = {2.0},
doi = {10.5281/zenodo.5525023},
url = {https://doi.org/10.5281/zenodo.5525023}
} Thorsten-Voice Dataset 2022.10 (Neutral)
- Number of recordings
- 12,432
- Audio duration
- 11 hours
- Sample rate
- 22,050 Hz
- Channels
- Mono
- Normalization
- -24 dB
- Speaking rate (avg)
- 17.5 characters/second
Citation (BibTeX)
@dataset{muller_thorsten_2022_7265581,
author = {Müller, Thorsten and Kreutz, Dominik},
title = {ThorstenVoice Dataset 2022.10},
month = oct,
year = 2022,
publisher = {Zenodo},
version = {1.0},
doi = {10.5281/zenodo.7265581},
url = {https://doi.org/10.5281/zenodo.7265581}
} Thorsten-Voice Dataset 2023.09 (Hessian dialect)
- Number of recordings
- 2,108
- Audio duration
- approx. 2 hours
- Sample rate
- 22,050 Hz
- Channels
- Mono
- Normalization
- -24 dB
Citation (BibTeX)
@dataset{muller_2024_10511260,
author = {Müller, Thorsten and Kreutz, Dominik},
title = {Thorsten-Voice Dataset 2023.09 Hessisch},
month = jan,
year = 2024,
publisher = {Zenodo},
doi = {10.5281/zenodo.10511260},
url = {https://doi.org/10.5281/zenodo.10511260}
} Thorsten-Voice Dataset (TV-44kHz-Full)
All recordings in a single dataset, in the original 44 kHz sample rate, logically split into subsets and enriched with metadata on duration, speaking rate, recording month and quality.
Citation (BibTeX)
@misc{thorsten_mueller_2024,
author = {Thorsten Müller},
title = {TV-44kHz-Full (Revision ff427ec)},
year = 2024,
url = {https://huggingface.co/datasets/Thorsten-Voice/TV-44kHz-Full},
doi = {10.57967/hf/3290},
publisher = {Hugging Face}
}