#138 · Primary category: Speech & Audio
MimikaStudio
MimikaStudio - A local-first application for macOS (Apple Silicon) + Agentic MCP Support
Project last updated:04/01/26
GitHub Stars
731
Forks
97
Contributors
1
License
GPL-3.0
Why we included this project
Voice cloning and TTS tools often force you to choose between a desktop app and a server you can script. MimikaStudio tries to be both, and it runs entirely on-device on Apple Silicon. You can clone a voice from a few seconds of reference audio with Qwen3-TTS or Chatterbox, or generate speech with the faster Kokoro and Supertonic models, all behind one interface. It also works as a document reader: it reads PDF, DOCX, EPUB, Markdown, and TXT files aloud with sentence-level highlighting, and can queue up a whole book as an audiobook using voice presets you can reuse. If you'd rather drive TTS from your own scripts, the same engine exposes an agentic server with a job queue and an API. One caveat: binaries are macOS-only for now, so Windows users will have to wait for the planned release.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
whisper.cpp
Port of OpenAI's Whisper model in C/C++
Real-Time-Voice-Cloning
Clone a voice in 5 seconds to generate arbitrary speech in real-time
VibeVoice
Open-Source Frontier Voice AI
voicebox
The open-source AI voice studio. Clone, dictate, create.
TTS
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production