Who is betting on what in AI, and against whom.
Researcher
Turing Award 2024
Co-author of the standard reinforcement-learning textbook and 2024 Turing laureate. Wrote The Bitter Lesson. Argues that intelligence is continual learning from a reward stream and that current deep learning cannot do it. Founded Oak Lab in Toronto in July 2026 after three years at Keen.
Known for Reinforcement Learning: An Introduction; The Bitter Lesson; Temporal-difference learning; The Alberta Plan
Experience / RL Intelligence comes from an agent's own stream of experience, not from a fixed corpus of human text.