#429 · Primary category: Education & Research
TinyEngram
Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.
Project last updated:05/21/26
GitHub Stars
1.3K
Forks
83
Contributors
2
License
Other
Why we included this project
If you're exploring ways to inject domain knowledge into large models without full retraining, TinyEngram offers a concrete alternative to LoRA. It's an open research implementation of DeepSeek's Engram architecture, an N-gram-style memory module with gated retrieval, applied to Qwen language models and extended to Stable Diffusion for concept injection. The repo includes training scripts, pinned dependencies, logs, and detailed experimental reports that back its central claims: Engram-based memory injection can be more parameter-efficient than LoRA while resisting catastrophic forgetting, and it composes cleanly across text and vision. For researchers and engineers working on memory-injection and fine-tuning techniques, this is a reproducible reference rather than a polished product, useful for understanding how a lightweight memory module can teach a model new subjects without retraining the backbone.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
prompts.chat
f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
JavaGuide
Java Interview & Backend General Interview Guide, covering computer fundamentals, databases, distributed systems, high concurrency, system design, and AI application development.
system-prompts-and-models-of-ai-tools
A curated collection of system prompts, internal tools, and AI models from popular AI assistants and coding agents.
30-seconds-of-code
Coding articles to level up your development skills
generative-ai-for-beginners
21 Lessons, Get Started Building with Generative AI