#238 · Primary category: AI Tool Directories & Curated Lists

ai-audio-datasets

aigc artificial-intelligence audio audio-effect audio-generation datasets deep-learning machine-learning music-generation

AI Audio Datasets (AI-ADS) 🎵, including Speech, Music, and Sound Effects, which can provide training data for Generative AI, AIGC, AI model training, intelligent audio tool development, and audio applications.

Project last updated:07/08/25

GitHub Stars

962

Forks

99

Contributors

10

License

MIT

Why we included this project

This is a searchable index of open audio datasets covering speech, music, and sound effects, useful for anyone training or fine-tuning audio models. Instead of digging through scattered forum threads and papers, you get a single curated list where each entry says what the corpus contains, roughly how large it is, and what it is typically used for. Teams working on text-to-speech, speaker recognition, audio captioning, voice conversion, or generative music can scan the list to shortlist candidates before spending bandwidth on downloads. It balances well-known corpora like Common Voice and AISHELL with specialized research releases, so it also works as an orientation guide for newcomers to the field. It is a reference list rather than software you run, but as a gateway to the underlying datasets it can save real time during the early stages of a project.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category