有时,需要向非理工科的或没有基础的朋友们解释语言模型。这些博客很能帮上忙:
- 故事开始于 2017:transformers.run
- LLM:The Illustrated GPT-2 (Visualizing Transformer Language Models) – Jay Alammar – Visualizing machine learning one concept at a time.
- KV cache:Transformers KV Caching Explained | by João Lages | Medium
- attention:Attention? Attention! | Lil’Log
- RL:Reinforcement Learning from Human Feedback (RLHF) - a simplified explanation,Medium
- 一个 CHI 2026 作品:transformer-explainer
优秀的博客作者: