← all repositories

alessiodm/drl-zh

An educational Jupyter Notebook course teaching deep reinforcement learning by building algorithms from scratch.

2.3k stars Jupyter Notebook LearningAgents
drl-zh
Not currently ranked — collecting fresh signals.
star history

The course starts with MDPs and tabular RL, progressing to DQN, REINFORCE, actor-critic methods, DDPG, TD3, SAC, and PPO. Advanced notebooks cover RLHF with PPO, DPO, and GRPO for language models, Decision Transformers, and world models like Dreamer. Students write code in guided TODO sections with complete solutions provided.

Frequently asked

What is alessiodm/drl-zh?
An educational Jupyter Notebook course teaching deep reinforcement learning by building algorithms from scratch.
Is drl-zh open source?
Yes — alessiodm/drl-zh is open source, released under the MIT license.
What language is drl-zh written in?
alessiodm/drl-zh is primarily written in Jupyter Notebook.
How popular is drl-zh?
alessiodm/drl-zh has 2.3k stars on GitHub.
Where can I find drl-zh?
alessiodm/drl-zh is on GitHub at https://github.com/alessiodm/drl-zh.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.