#207 · Primary category: Education & Research

LLM-RL-Visualized

ai algorithm deep-learning llm machine-learning natural-language-processing nlp-machine-learning reinforcement-learning transformers vlm

🌟100+ 原创 LLM / RL 原理图📚,《大模型算法》作者巨献!💥(100+ LLM/RL Algorithm Maps )

Project last updated:08/29/26

GitHub Stars

4.8K

Forks

464

Contributors

1

License

Other

Why we included this project

This repo collects more than a hundred original, hand-drawn diagrams that walk through how large language models and reinforcement learning algorithms work, from transformer internals and attention to SFT, LoRA, DPO, RLHF, PPO, GRPO, and reasoning techniques like CoT and MCTS. It is a study resource, not runnable software, so the value is for anyone trying to understand or teach these systems: the path from MDP fundamentals to modern policy-gradient training is mapped out in one browsable set of SVGs that stay crisp at any zoom. That makes it useful for developers brushing up before a project, students prepping for interviews, or anyone who has to explain RLHF trade-offs to a team without hand-waving. A companion book adds deeper written explanations, and the author keeps correcting and expanding the charts, so the collection tracks newer methods such as GRPO and VLM architectures.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category