【极致中配】UC Berkery CS 285: 深度强化学习 Deep RL, 2023
伯克利CS 285研究生课程,系统讲授深度强化学习:从模仿学习、策略梯度、Q学习到基于模型、离线RL、控制即推理、逆强化学习及序列模型前沿。
- 难度
- 难度 5/5 — 研究生级深度强化学习课程,理论推导密集,需扎实机器学习基础
- 适合人群
- 有深度学习基础、想系统学习强化学习的研究生或研究者
- 前置要求
- 机器学习(监督学习、损失函数与优化基础)、深度学习(神经网络、反向传播、PyTorch/TensorFlow)、概率论与统计(随机变量、期望、分布)、微积分与线性代数(梯度、矩阵运算)
- 课程规模
- 99 讲 · 3459播放
主题覆盖
模仿学习策略梯度Actor-Critic值函数与Q学习基于模型的RL离线强化学习探索变分推断控制即推理逆强化学习元学习与迁移RL与语言模型
课程大纲(99 讲)
- P1 · CS 285: Lecture 1, Introduction. Part 16 分钟
- P2 · CS 285: Lecture 1, Introduction. Part 212 分钟
- P3 · CS 285: Lecture 1, Introduction. Part 320 分钟
- P4 · CS 285: Lecture 2, Imitation Learning. Part 115 分钟
- P5 · CS 285: Lecture 2, Imitation Learning. Part 215 分钟
- P6 · CS 285: Lecture 2, Imitation Learning. Part 321 分钟
- P7 · CS 285: Lecture 2, Imitation Learning. Part 45 分钟
- P8 · CS 285: Lecture 2, Imitation Learning. Part 56 分钟
- P9 · CS 285: Lecture 4, Part 118 分钟
- P10 · CS 285: Lecture 4, Part 24 分钟
- P11 · CS 285: Lecture 4, Part 36 分钟
- P12 · CS 285: Lecture 4, Part 44 分钟
- P13 · CS 285: Lecture 4, Part 56 分钟
- P14 · CS 285: Lecture 4, Part 62 分钟
- P15 · CS 285: Lecture 5, Part 110 分钟
- P16 · CS 285: Lecture 5, Part 29 分钟
- P17 · CS 285: Lecture 5, Part 310 分钟
- P18 · CS 285: Lecture 5, Part 411 分钟
- P19 · CS 285: Lecture 5, Part 54 分钟
- P20 · CS 285: Lecture 5, Part 69 分钟
- P21 · CS 285: Lecture 6, Part 117 分钟
- P22 · CS 285: Lecture 6, Part 212 分钟
- P23 · CS 285: Lecture 6, Part 313 分钟
- P24 · CS 285: Lecture 6, Part 411 分钟
- P25 · CS 285: Lecture 6, Part 52 分钟
- P26 · CS 285: Lecture 7, Part 111 分钟
- P27 · CS 285: Lecture 7, Part 210 分钟
- P28 · CS 285: Lecture 7, Part 38 分钟
- P29 · CS 285: Lecture 7, Part 412 分钟
- P30 · CS 285: Lecture 8, Part 110 分钟
- P31 · CS 285: Lecture 8, Part 27 分钟
- P32 · CS 285: Lecture 8, Part 35 分钟
- P33 · CS 285: Lecture 8, Part 415 分钟
- P34 · CS 285: Lecture 8, Part 56 分钟
- P35 · CS 285: Lecture 8, Part 67 分钟
- P36 · CS 285: Lecture 9, Part 114 分钟
- P37 · CS 285: Lecture 9, Part 214 分钟
- P38 · CS 285: Lecture 9, Part 34 分钟
- P39 · CS 285: Lecture 9, Part 415 分钟
- P40 · CS 285: Lecture 10, Part 112 分钟
- P41 · CS 285: Lecture 10, Part 216 分钟
- P42 · CS 285: Lecture 10, Part 316 分钟
- P43 · CS 285: Lecture 10, Part 49 分钟
- P44 · CS 285: Lecture 10, Part 54 分钟
- P45 · CS 285: Lecture 11, Part 112 分钟
- P46 · CS 285: Lecture 11, Part 26 分钟
- P47 · CS 285: Lecture 11, Part 312 分钟
- P48 · CS 285: Lecture 11, Part 45 分钟
- P49 · CS 285: Lecture 11, Part 512 分钟
- P50 · CS 285: Lecture 12, Part 1: Model-Based RL with Policies10 分钟
- P51 · CS 285: Lecture 12, Part 2: Model-Based RL with Policies10 分钟
- P52 · CS 285: Lecture 12, Part 3: Model-Based RL with Policies8 分钟
- P53 · CS 285: Lecture 12, Part 4: Model-Based RL with Policies20 分钟
- P54 · CS 285: Lecture 13, Part 114 分钟
- P55 · CS 285: Lecture 13, Part 211 分钟
- P56 · CS 285: Lecture 13, Part 39 分钟
- P57 · CS 285: Lecture 13, Part 49 分钟
- P58 · CS 285: Lecture 13, Part 55 分钟
- P59 · CS 285: Lecture 13, Part 69 分钟
- P60 · CS 285: Lecture 14, Part 19 分钟
- P61 · CS 285: Lecture 14, Part 211 分钟
- P62 · CS 285: Lecture 14, Part 310 分钟
- P63 · CS 285: Lecture 14, Part 45 分钟
- P64 · CS 285: Lecture 15, Part 1: Offline Reinforcement Learning26 分钟
- P65 · CS 285: Lecture 15, Part 2: Offline Reinforcement Learning17 分钟
- P66 · CS 285: Lecture 15, Part 3: Offline Reinforcement Learning15 分钟
- P67 · CS 285: Lecture 16, Part 1: Offline Reinforcement Learning 222 分钟
- P68 · CS 285: Lecture 16, Part 2: Offline Reinforcement Learning 25 分钟
- P69 · CS 285: Lecture 16, Part 3: Offline Reinforcement Learning 213 分钟
- P70 · CS 285: Lecture 16, Part 4: Offline Reinforcement Learning 29 分钟
- P71 · CS 285: Lecture 17, Part 1: RL Theory35 分钟
- P72 · CS 285: Lecture 17, Part 2: RL Theory16 分钟
- P73 · CS 285: Lecture 18, Variational Inference, Part 114 分钟
- P74 · CS 285: Lecture 18, Variational Inference, Part 212 分钟
- P75 · CS 285: Lecture 18, Variational Inference, Part 312 分钟
- P76 · CS 285: Lecture 18, Variational Inference, Part 418 分钟
- P77 · CS 285: Lecture 19, Control as Inference, Part 114 分钟
- P78 · CS 285: Lecture 19, Control as Inference, Part 221 分钟
- P79 · CS 285: Lecture 19, Control as Inference, Part 314 分钟
- P80 · CS 285: Lecture 19, Control as Inference, Part 47 分钟
- P81 · CS 285: Lecture 19, Control as Inference, Part 57 分钟
- P82 · CS 285: Lecture 20, Inverse Reinforcement Learning, Part 116 分钟
- P83 · CS 285: Lecture 20, Inverse Reinforcement Learning, Part 28 分钟
- P84 · CS 285: Lecture 20, Inverse Reinforcement Learning, Part 35 分钟
- P85 · CS 285: Lecture 20, Inverse Reinforcement Learning, Part 410 分钟
- P86 · CS 285: Eric Mitchell: Reinforcement Learning from Human Feedback: Algorithms &36 分钟
- P87 · CS 285: Andrea Zanette: Towards a Statistical Foundation for Reinforcement Learn37 分钟
- P88 · CS 285: Lecture 21, RL with Sequence Models & Language Models, Part 119 分钟
- P89 · CS 285: Lecture 21, RL with Sequence Models & Language Models, Part 216 分钟
- P90 · CS 285: Lecture 21, RL with Sequence Models & Language Models, Part 311 分钟
- P91 · CS 285: Lecture 22, Part 1: Transfer Learning & Meta-Learning29 分钟
- P92 · CS 285: Lecture 22, Part 2: Transfer Learning & Meta-Learning6 分钟
- P93 · CS 285: Lecture 22, Part 3: Transfer Learning & Meta-Learning8 分钟
- P94 · CS 285: Lecture 22, Part 4: Transfer Learning & Meta-Learning5 分钟
- P95 · CS 285: Lecture 22, Part 5: Transfer Learning & Meta-Learning9 分钟
- P96 · CS 285: Lecture 23, Part 1: Challenges & Open Problems19 分钟
- P97 · CS 285: Lecture 23, Part 2: Challenges & Open Problems26 分钟
- P98 · CS 285: Guest Lecture: Aviral Kumar37 分钟
- P99 · CS 285: Guest Lecture: Dorsa Sadigh43 分钟
本课程卡由 AI 生成,可能存在误差,欢迎反馈。