强化学习课程进阶EN4–8 周★ 5.0k114 章可站内阅读
HF Deep RL Class
Hugging Face 深度强化学习课程
Hugging Face· 4,981 stars
Hugging Face 深度强化学习课程仓库,按单元组织 Q-Learning 到 Actor-Critic 等主题,配合 notebook 跟练。适合想把 RL 理论落到可运行实验的工程师。
为什么收录 · 与 Spinning Up、Hands-on RL 互补,补齐 HF 官方深度 RL 课程轨。
强化学习Hugging FaceDQN策略梯度
这本书强在哪
- GitHub 高星真学习内容
- 可导入站内阅读
- 材料结构清晰
- 适合系统跟学
建议怎么学
- 01按大纲顺序推进
- 02每章留下自己的实验记录
- 03卡点时回看对应知识库文章
适合谁 / 前置
- Python 基础
- 愿意跟练代码或笔记
学完得到什么
- 掌握该教程主线知识点
- 能复现关键实验或练习
- 建立可对照的学习笔记
- 为下一阶段课程打底
目录
来自站内阅读器镜像;点击章节直接阅读
第 0 单元3 章
第 1 单元13 章
- 04Additional Readingsmarkdown
- 05Conclusionmarkdown
- 06The “Deep” in Reinforcement Learningmarkdown
- 07The Exploration/Exploitation trade-offmarkdown
- 08Glossarymarkdown
- 09Train your first Deep Reinforcement Learning Agent 🤖markdown
- 10Introduction to Deep Reinforcement Learningmarkdown
- 11Quizmarkdown
另有 5 章 · 进入站内阅读查看完整目录
第 2 单元15 章
- 17Additional Readingsmarkdown
- 18The Bellman Equation: simplify our value estimationmarkdown
- 19Conclusionmarkdown
- 20Glossarymarkdown
- 21Hands-onmarkdown
- 22Introduction to Q-Learningmarkdown
- 23Monte Carlo vs Temporal Difference Learningmarkdown
- 24Mid-way Quizmarkdown
另有 7 章 · 进入站内阅读查看完整目录
第 3 单元9 章
- 32Additional Readingsmarkdown
- 33Conclusionmarkdown
- 34The Deep Q-Learning Algorithmmarkdown
- 35The Deep Q-Network (DQN)markdown
- 36From Q-Learning to Deep Q-Learningmarkdown
- 37Glossarymarkdown
- 38Hands-onmarkdown
- 39Deep Q-Learningmarkdown
另有 1 章 · 进入站内阅读查看完整目录
第 4 单元10 章
- 41Additional Readingsmarkdown
- 42The advantages and disadvantages of policy-gradient methodsmarkdown
- 43Conclusionmarkdown
- 44Glossarymarkdown
- 45Hands onmarkdown
- 46Introductionmarkdown
- 47(Optional) the Policy Gradient Theoremmarkdown
- 48Diving deeper into policy-gradient methodsmarkdown
另有 2 章 · 进入站内阅读查看完整目录
第 5 单元9 章
- 51Bonus: Learn to create your own environments with Unity and MLAgentsmarkdown
- 52Conclusionmarkdown
- 53(Optional) What is Curiosity in Deep Reinforcement Learning?markdown
- 54Hands-onmarkdown
- 55How do Unity ML-Agents work?markdown
- 56An Introduction to Unity ML-Agentsmarkdown
- 57The Pyramid environmentmarkdown
- 58Quizmarkdown
另有 1 章 · 进入站内阅读查看完整目录
第 6 单元7 章
第 7 单元8 章
- 67Additional Readingsmarkdown
- 68Conclusionmarkdown
- 69Hands-onmarkdown
- 70An introduction to Multi-Agents Reinforcement Learning (MARL)markdown
- 71Introductionmarkdown
- 72Designing Multi-Agents systemsmarkdown
- 73Quizmarkdown
- 74Self-Play: a classic technique to train competitive agents in adversarial gamesmarkdown
第 8 单元10 章
- 75Additional Readingsmarkdown
- 76Introducing the Clipped Surrogate Objective Functionmarkdown
- 77Conclusionmarkdown
- 78Conclusionmarkdown
- 79Hands-onmarkdown
- 80Hands-on: advanced Deep Reinforcement Learning. Using Sample Factory to play Doom from pixelsmarkdown
- 81Introduction to PPO with Sample-Factorymarkdown
- 82Introductionmarkdown
另有 2 章 · 进入站内阅读查看完整目录
其他30 章
- 85The certification processmarkdown
- 86Congratulationsmarkdown
- 87Live 1: How the course work, Q&A, and playing with Huggymarkdown
- 88Conclusionmarkdown
- 89How Huggy worksmarkdown
- 90Introductionmarkdown
- 91Play with Huggymarkdown
- 92Let's train and play with Huggy 🐶markdown
另有 22 章 · 进入站内阅读查看完整目录
策展亮点章节
Q-LearningDeep Q-LearningPolicy GradientActor-CriticMulti-AgentOffline RL
仓库数据
Stars4,981
Forks808
主要语言MDX
创建2022/4/21
更新2026/8/11
ad165751(main)· 许可证 Apache License 2.0 (Apache-2.0)。 上游更新后可通过同步脚本刷新。