AI 学者图谱Ai Scholar Graph ↗ Richard Sutton

Richard SuttonRichard Sutton

University of Alberta

Father of Reinforcement Learning · Turing Award 2024Father of Reinforcement Learning · Turing Award 2024

rlagents
Born 1956 in the USA. Earned a BA in Psychology from Stanford (1978), then a PhD in Computer Science from the University of Massachusetts Amherst (1984) under Andrew Barto -- together they founded the field of reinforcement learning. Joined the University of Alberta (2003) where he became a Distinguished Research Professor and iCORE Chair. Co-authored 'Reinforcement Learning: An Introduction' (1998, 2nd ed. 2018) with Barto -- THE standard textbook that trained virtually every RL researcher alive. Invented temporal-difference (TD) learning, policy gradient methods, and the Dyna architecture. Served as Distinguished Research Scientist at DeepMind Alberta. His 2019 essay 'The Bitter Lesson' -- arguing that general methods leveraging computation always win over hand-crafted approaches -- became one of the most cited philosophical pieces in modern AI. Shared the 2024 ACM Turing Award with Barto. Remains RL's contrarian conscience: argues LLMs lack true world understanding, that intelligence must be grounded in runtime experience ('The Era of Experience', 2025, with David Silver), and keeps pursuing a simple general agent architecture (the Alberta Plan).
Born 1956 in the USA. Earned a BA in Psychology from Stanford (1978), then a PhD in Computer Science from the University of Massachusetts Amherst (1984) under Andrew Barto -- together they founded the field of reinforcement learning. Joined the University of Alberta (2003) where he became a Distinguished Research Professor and iCORE Chair. Co-authored 'Reinforcement Learning: An Introduction' (1998, 2nd ed. 2018) with Barto -- THE standard textbook that trained virtually every RL researcher alive. Invented temporal-difference (TD) learning, policy gradient methods, and the Dyna architecture. Served as Distinguished Research Scientist at DeepMind Alberta. His 2019 essay 'The Bitter Lesson' -- arguing that general methods leveraging computation always win over hand-crafted approaches -- became one of the most cited philosophical pieces in modern AI. Shared the 2024 ACM Turing Award with Barto. Remains RL's contrarian conscience: argues LLMs lack true world understanding, that intelligence must be grounded in runtime experience ('The Era of Experience', 2025, with David Silver), and keeps pursuing a simple general agent architecture (the Alberta Plan).

同一人物 · 姊妹站:听 TA 的播客访谈(AI Podcast) · 读 TA 的论文与长文(AI Paper)

在关系图谱中查看 Richard Sutton →View Richard Sutton in the graph → 听 Richard Sutton 的 AI 播客 →Listen on AI Podcast → 读 Richard Sutton 的论文与长文 →Read on AI Paper →

时间线Timeline

关系网络Connections

David SilverInfluenced RL foundations, AlbertaDemis HassabisDeepMind Alberta, RL foundations ← 返回完整图谱← Back to the graph