Anthropic 研究员 · 规模化与强化学习Researcher at Anthropic · Scaling & RL
deep-learningrl
澳大利亚 AI 研究者,在 Anthropic 从事大模型推理、强化学习与规模化工程。此前在 Google DeepMind 参与 Gemini 模型研发,后加入 Anthropic。他以清晰的公开讲解著称(包括多次长篇访谈),阐释 RL、推理与规模化法则如何决定前沿模型的走向,以及通往更强能力 Agent 的路径。
Australian AI researcher working at Anthropic on large language model reasoning, reinforcement learning, and the engineering of scaling. He previously worked at Google DeepMind on the Gemini models before joining Anthropic. Known for clear public explanations, including long-form interviews, of how RL, inference and scaling laws shape the trajectory of frontier models and progress toward more capable agents.