#164 · Primary category: LLM Application Frameworks
motorhead
🧠 Motorhead is a memory and information retrieval server for LLMs.
Project last updated:07/22/25
GitHub Stars
917
Forks
87
Contributors
12
License
Apache-2.0
Why we included this project
Most teams building chat applications on top of LLMs end up reimplementing conversation memory by hand, and this project exists to save that repeated work. It runs as a small standalone server that stores and returns the recent message history for a chat session, letting you pull up what the model has already said and feed it back when the next request comes in. Because it is written in Rust and keeps the memory handling out of your application code, your service can stay focused on generating responses instead of juggling context. One caveat worth knowing up front: the project is officially deprecated and no longer maintained, so view it as a reference implementation or a starting point rather than something to rely on long-term.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
langchain
The agent engineering platform.
dify
Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without rebuilding the stack.
headroom
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
litellm
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]
llama_index
LlamaIndex is the leading document agent and OCR platform