#558 · Primary category: Education & Research
dla
Deep learning for audio processing
Project last updated:12/15/25
GitHub Stars
762
Forks
123
Contributors
14
License
MIT
Why we included this project
This is the full teaching repository for a university course on deep learning for audio, run at HSE's CS faculty. Instead of a single tool or model, it walks through a structured weekly progression from digital signal processing fundamentals through automatic speech recognition, source separation, audio-visual learning, and text-to-speech, with lectures, seminar notebooks, and homework for each stage. Developers and researchers working on speech and audio ML will find ready-made PyTorch pipelines, experiment-tracking setups, and practical code for CTC and RNN-T training, beam search, and separation architectures like Demucs and ConvTasNet. The course is maintained and re-run each year, so the materials track current practice; that makes them a good self-study path or a base for building your own team onboarding curriculum.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
prompts.chat
f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
JavaGuide
Java Interview & Backend General Interview Guide, covering computer fundamentals, databases, distributed systems, high concurrency, system design, and AI application development.
system-prompts-and-models-of-ai-tools
A curated collection of system prompts, internal tools, and AI models from popular AI assistants and coding agents.
30-seconds-of-code
Coding articles to level up your development skills
generative-ai-for-beginners
21 Lessons, Get Started Building with Generative AI