#154 · Primary category: Inference & Local Deploy
LLM-Hub
Local LLM, image&video&music generator, vibecode like cursor with local models on your phone
Project last updated:08/24/26
GitHub Stars
572
Forks
115
Contributors
6
License
MIT
Why we included this project
LLM Hub keeps the models on your phone instead of sending prompts to a cloud API. Chat, image, video, and music generation all run locally, with CPU/GPU/NPU acceleration and support for multiple model formats so you can pick weights that fit your hardware. It's a native Android and iOS app, and beyond plain chat it adds a vibe-coding mode that turns an app idea into a live HTML preview, an agent that can call device and MCP tools, and RAG-backed memory with web search. For teams building privacy-sensitive mobile experiences, or hobbyists who want a self-contained local assistant, it's a useful reference for what on-device inference can do today.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
ollama
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
llama.cpp
LLM inference in C/C++
vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
gpt4all
GPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use.
LocalAI
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.