#263 · Primary category: AI Tool Directories & Curated Lists

awesome-speech-recognition-speech-synthesis-papers

acoustic-model attention-mechanism automatic-speech-recognition cnn diffusion-models dnn language-model neural-network papers recognition-synthesis rnn roadmap seq2seq singing-voice-synthesis speaker-verification speech-recognition speech-synthesis timit-dataset tts voice-conversion

Automatic Speech Recognition (ASR), Speaker Verification, Speech Synthesis, Text-to-Speech (TTS), Language Modelling, Singing Voice Synthesis (SVS), Voice Conversion (VC)

Project last updated:10/19/23

GitHub Stars

3.1K

Forks

514

Contributors

12

License

MIT

Why we included this project

This index of research papers covers the speech and audio field, from classic hidden Markov model and finite-state transducer work to the attention-based sequence-to-sequence and diffusion models used today. If you're new to speech technology, the organization helps you follow how modern ASR and TTS systems evolved instead of dropping you into random papers. For practitioners scoping a project, it works as a reading map: papers are grouped by task (recognition, verification, synthesis, voice conversion, language and music modelling) and each entry links straight to its PDF, so you don't have to dig through arXiv and conference archives. Since it's a bibliography rather than runnable code, treat it as a study companion and a starting point for your own literature review, not as something you deploy.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category