#42 · Primary category: Speech & Audio
abogen
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
Project last updated:08/29/26
GitHub Stars
5.8K
Forks
432
Contributors
19
License
MIT
Why we included this project
Anyone with a shelf of EPUBs or PDFs they keep meaning to read can point Abogen at those files and get natural-sounding speech plus matching captions to follow along with. It is built on the lightweight Kokoro-82M text-to-speech model, and the speed shows: the project's own demo renders about a minute of audio in roughly five seconds, with subtitles synced to the voice. That makes it handy both for turning personal ebook collections into audiobooks and for creators who need quick voiceover with captions for short social videos. It runs as a command-line tool, so it slots into scripts and batch workflows without requiring heavy infrastructure.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
whisper.cpp
Port of OpenAI's Whisper model in C/C++
Real-Time-Voice-Cloning
Clone a voice in 5 seconds to generate arbitrary speech in real-time
VibeVoice
Open-Source Frontier Voice AI
voicebox
The open-source AI voice studio. Clone, dictate, create.
TTS
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production