跳到主要内容
步芽

【极致中配】UC Berkery CS 285: 深度强化学习 Deep RL, 2023

伯克利CS 285研究生课程,系统讲授深度强化学习:从模仿学习、策略梯度、Q学习到基于模型、离线RL、控制即推理、逆强化学习及序列模型前沿。

难度
难度 5/5研究生级深度强化学习课程,理论推导密集,需扎实机器学习基础
适合人群
有深度学习基础、想系统学习强化学习的研究生或研究者
前置要求
机器学习(监督学习、损失函数与优化基础)、深度学习(神经网络、反向传播、PyTorch/TensorFlow)、概率论与统计(随机变量、期望、分布)、微积分与线性代数(梯度、矩阵运算)
课程规模
99 · 3459播放

主题覆盖

模仿学习策略梯度Actor-Critic值函数与Q学习基于模型的RL离线强化学习探索变分推断控制即推理逆强化学习元学习与迁移RL与语言模型

课程大纲(99 讲)

  1. P1 · CS 285: Lecture 1, Introduction. Part 16 分钟
  2. P2 · CS 285: Lecture 1, Introduction. Part 212 分钟
  3. P3 · CS 285: Lecture 1, Introduction. Part 320 分钟
  4. P4 · CS 285: Lecture 2, Imitation Learning. Part 115 分钟
  5. P5 · CS 285: Lecture 2, Imitation Learning. Part 215 分钟
  6. P6 · CS 285: Lecture 2, Imitation Learning. Part 321 分钟
  7. P7 · CS 285: Lecture 2, Imitation Learning. Part 45 分钟
  8. P8 · CS 285: Lecture 2, Imitation Learning. Part 56 分钟
  9. P9 · CS 285: Lecture 4, Part 118 分钟
  10. P10 · CS 285: Lecture 4, Part 24 分钟
  11. P11 · CS 285: Lecture 4, Part 36 分钟
  12. P12 · CS 285: Lecture 4, Part 44 分钟
  13. P13 · CS 285: Lecture 4, Part 56 分钟
  14. P14 · CS 285: Lecture 4, Part 62 分钟
  15. P15 · CS 285: Lecture 5, Part 110 分钟
  16. P16 · CS 285: Lecture 5, Part 29 分钟
  17. P17 · CS 285: Lecture 5, Part 310 分钟
  18. P18 · CS 285: Lecture 5, Part 411 分钟
  19. P19 · CS 285: Lecture 5, Part 54 分钟
  20. P20 · CS 285: Lecture 5, Part 69 分钟
  21. P21 · CS 285: Lecture 6, Part 117 分钟
  22. P22 · CS 285: Lecture 6, Part 212 分钟
  23. P23 · CS 285: Lecture 6, Part 313 分钟
  24. P24 · CS 285: Lecture 6, Part 411 分钟
  25. P25 · CS 285: Lecture 6, Part 52 分钟
  26. P26 · CS 285: Lecture 7, Part 111 分钟
  27. P27 · CS 285: Lecture 7, Part 210 分钟
  28. P28 · CS 285: Lecture 7, Part 38 分钟
  29. P29 · CS 285: Lecture 7, Part 412 分钟
  30. P30 · CS 285: Lecture 8, Part 110 分钟
  31. P31 · CS 285: Lecture 8, Part 27 分钟
  32. P32 · CS 285: Lecture 8, Part 35 分钟
  33. P33 · CS 285: Lecture 8, Part 415 分钟
  34. P34 · CS 285: Lecture 8, Part 56 分钟
  35. P35 · CS 285: Lecture 8, Part 67 分钟
  36. P36 · CS 285: Lecture 9, Part 114 分钟
  37. P37 · CS 285: Lecture 9, Part 214 分钟
  38. P38 · CS 285: Lecture 9, Part 34 分钟
  39. P39 · CS 285: Lecture 9, Part 415 分钟
  40. P40 · CS 285: Lecture 10, Part 112 分钟
  41. P41 · CS 285: Lecture 10, Part 216 分钟
  42. P42 · CS 285: Lecture 10, Part 316 分钟
  43. P43 · CS 285: Lecture 10, Part 49 分钟
  44. P44 · CS 285: Lecture 10, Part 54 分钟
  45. P45 · CS 285: Lecture 11, Part 112 分钟
  46. P46 · CS 285: Lecture 11, Part 26 分钟
  47. P47 · CS 285: Lecture 11, Part 312 分钟
  48. P48 · CS 285: Lecture 11, Part 45 分钟
  49. P49 · CS 285: Lecture 11, Part 512 分钟
  50. P50 · CS 285: Lecture 12, Part 1: Model-Based RL with Policies10 分钟
  51. P51 · CS 285: Lecture 12, Part 2: Model-Based RL with Policies10 分钟
  52. P52 · CS 285: Lecture 12, Part 3: Model-Based RL with Policies8 分钟
  53. P53 · CS 285: Lecture 12, Part 4: Model-Based RL with Policies20 分钟
  54. P54 · CS 285: Lecture 13, Part 114 分钟
  55. P55 · CS 285: Lecture 13, Part 211 分钟
  56. P56 · CS 285: Lecture 13, Part 39 分钟
  57. P57 · CS 285: Lecture 13, Part 49 分钟
  58. P58 · CS 285: Lecture 13, Part 55 分钟
  59. P59 · CS 285: Lecture 13, Part 69 分钟
  60. P60 · CS 285: Lecture 14, Part 19 分钟
  61. P61 · CS 285: Lecture 14, Part 211 分钟
  62. P62 · CS 285: Lecture 14, Part 310 分钟
  63. P63 · CS 285: Lecture 14, Part 45 分钟
  64. P64 · CS 285: Lecture 15, Part 1: Offline Reinforcement Learning26 分钟
  65. P65 · CS 285: Lecture 15, Part 2: Offline Reinforcement Learning17 分钟
  66. P66 · CS 285: Lecture 15, Part 3: Offline Reinforcement Learning15 分钟
  67. P67 · CS 285: Lecture 16, Part 1: Offline Reinforcement Learning 222 分钟
  68. P68 · CS 285: Lecture 16, Part 2: Offline Reinforcement Learning 25 分钟
  69. P69 · CS 285: Lecture 16, Part 3: Offline Reinforcement Learning 213 分钟
  70. P70 · CS 285: Lecture 16, Part 4: Offline Reinforcement Learning 29 分钟
  71. P71 · CS 285: Lecture 17, Part 1: RL Theory35 分钟
  72. P72 · CS 285: Lecture 17, Part 2: RL Theory16 分钟
  73. P73 · CS 285: Lecture 18, Variational Inference, Part 114 分钟
  74. P74 · CS 285: Lecture 18, Variational Inference, Part 212 分钟
  75. P75 · CS 285: Lecture 18, Variational Inference, Part 312 分钟
  76. P76 · CS 285: Lecture 18, Variational Inference, Part 418 分钟
  77. P77 · CS 285: Lecture 19, Control as Inference, Part 114 分钟
  78. P78 · CS 285: Lecture 19, Control as Inference, Part 221 分钟
  79. P79 · CS 285: Lecture 19, Control as Inference, Part 314 分钟
  80. P80 · CS 285: Lecture 19, Control as Inference, Part 47 分钟
  81. P81 · CS 285: Lecture 19, Control as Inference, Part 57 分钟
  82. P82 · CS 285: Lecture 20, Inverse Reinforcement Learning, Part 116 分钟
  83. P83 · CS 285: Lecture 20, Inverse Reinforcement Learning, Part 28 分钟
  84. P84 · CS 285: Lecture 20, Inverse Reinforcement Learning, Part 35 分钟
  85. P85 · CS 285: Lecture 20, Inverse Reinforcement Learning, Part 410 分钟
  86. P86 · CS 285: Eric Mitchell: Reinforcement Learning from Human Feedback: Algorithms &36 分钟
  87. P87 · CS 285: Andrea Zanette: Towards a Statistical Foundation for Reinforcement Learn37 分钟
  88. P88 · CS 285: Lecture 21, RL with Sequence Models & Language Models, Part 119 分钟
  89. P89 · CS 285: Lecture 21, RL with Sequence Models & Language Models, Part 216 分钟
  90. P90 · CS 285: Lecture 21, RL with Sequence Models & Language Models, Part 311 分钟
  91. P91 · CS 285: Lecture 22, Part 1: Transfer Learning & Meta-Learning29 分钟
  92. P92 · CS 285: Lecture 22, Part 2: Transfer Learning & Meta-Learning6 分钟
  93. P93 · CS 285: Lecture 22, Part 3: Transfer Learning & Meta-Learning8 分钟
  94. P94 · CS 285: Lecture 22, Part 4: Transfer Learning & Meta-Learning5 分钟
  95. P95 · CS 285: Lecture 22, Part 5: Transfer Learning & Meta-Learning9 分钟
  96. P96 · CS 285: Lecture 23, Part 1: Challenges & Open Problems19 分钟
  97. P97 · CS 285: Lecture 23, Part 2: Challenges & Open Problems26 分钟
  98. P98 · CS 285: Guest Lecture: Aviral Kumar37 分钟
  99. P99 · CS 285: Guest Lecture: Dorsa Sadigh43 分钟

本课程卡由 AI 生成,可能存在误差,欢迎反馈。