David Silver
4 entries between January 2016 and December 2020, 1 turning point.
Turning points
-
AlphaGo Zero learns Go from self-play alone
DeepMind announced AlphaGo Zero, trained only by playing itself from random play, with no human games and no hand-crafted features. After three days it beat the version that had defeated Lee Sedol, 100 games to nil. Nature published the work the next day.
Every entry
-
DeepMind’s MuZero plays Atari, Go, chess and shogi without rules
Nature published “Mastering Atari, Go, chess and shogi by planning with a learned model” on 23 December 2020. Julian Schrittwieser and colleagues at DeepMind described MuZero, which planned using a model it learned rather than rules it was given.
-
AlphaZero learns chess and shogi from the rules alone
DeepMind posted a paper describing AlphaZero, one algorithm that, given only the rules, reached superhuman play at chess, shogi and Go within 24 hours of self-play training and beat the reigning computer champion program in each game.
-
AlphaGo Zero learns Go from self-play alone
DeepMind announced AlphaGo Zero, trained only by playing itself from random play, with no human games and no hand-crafted features. After three days it beat the version that had defeated Lee Sedol, 100 games to nil. Nature published the work the next day.
-
Nature publishes the paper behind AlphaGo
Nature published “Mastering the game of Go with deep neural networks and tree search” by David Silver, Demis Hassabis and 18 co-authors at Google DeepMind. It disclosed that AlphaGo had already beaten the European champion Fan Hui five games to nil in October 2015.