AI 学者图谱AI Scholar Graph ↗
Richard SuttonRichard Sutton
University of Alberta
强化学习之父Father of Reinforcement Learning
rl
1956年生于美国。斯坦福心理学学士(1978),马萨诸塞大学计算机科学博士(1984,导师 Andrew Barto,两人共同建立了强化学习领域)。2003年加入阿尔伯塔大学。与 Barto 合著了《强化学习导论》(1998),是培养了几乎所有在世 RL 研究者的标准教科书。发明了时间差分(TD)学习、策略梯度方法和 Dyna 架构。曾任 DeepMind Alberta 杰出研究科学家。2019年发表的《The Bitter Lesson》成为现代 AI 最常被引用的哲学文章之一。
Born 1956 in the USA. Earned a BA in Psychology from Stanford (1978), then a PhD in Computer Science from the University of Massachusetts Amherst (1984) under Andrew Barto -- together they founded the field of reinforcement learning. Joined the University of Alberta (2003) where he became a Distinguished Research Professor and iCORE Chair. Co-authored 'Reinforcement Learning: An Introduction' (1998, 2nd ed. 2018) with Barto -- THE standard textbook that trained virtually every RL researcher alive. Invented temporal-difference (TD) learning, policy gradient methods, and the Dyna architecture. Served as Distinguished Research Scientist at DeepMind Alberta. His 2019 essay 'The Bitter Lesson' -- arguing that general methods leveraging computation always win over hand-crafted approaches -- became one of the most cited philosophical pieces in modern AI.
同一人物 · 姊妹站:听 TA 的播客访谈(AI Podcast) · 读 TA 的论文与长文(AI Paper)
在关系图谱中查看 Richard Sutton →View Richard Sutton in the graph →
听 Richard Sutton 的 AI 播客 →Listen to Richard Sutton on AI Podcast →
时间线Timeline
关系网络Connections
David SilverRL 理论奠基,阿尔伯塔Demis HassabisDeepMind Alberta,RL 基础
← 返回完整图谱← Back to the graph