#154 · Primary category: Inference & Local Deploy

LLM-Hub

ai gemma gemma4 gemma4-agent-skills gptoss granite imagegeneration lfm25 llama llm llm-inference mistral music musicgeneration phi4 rag stable-diffusion videogeneration whisper

Local LLM, image&video&music generator, vibecode like cursor with local models on your phone

Project last updated:08/24/26

GitHub Stars

572

Forks

115

Contributors

6

License

MIT

Why we included this project

LLM Hub keeps the models on your phone instead of sending prompts to a cloud API. Chat, image, video, and music generation all run locally, with CPU/GPU/NPU acceleration and support for multiple model formats so you can pick weights that fit your hardware. It's a native Android and iOS app, and beyond plain chat it adds a vibe-coding mode that turns an app idea into a live HTML preview, an agent that can call device and MCP tools, and RAG-backed memory with web search. For teams building privacy-sensitive mobile experiences, or hobbyists who want a self-contained local assistant, it's a useful reference for what on-device inference can do today.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category