David Silver

4 entries between January 2016 and December 2020, 1 turning point.

Turning points

  1. AlphaGo Zero learns Go from self-play alone

    DeepMind announced AlphaGo Zero, trained only by playing itself from random play, with no human games and no hand-crafted features. After three days it beat the version that had defeated Lee Sedol, 100 games to nil. Nature published the work the next day.

Every entry

  1. DeepMind’s MuZero plays Atari, Go, chess and shogi without rules

    Nature published “Mastering Atari, Go, chess and shogi by planning with a learned model” on 23 December 2020. Julian Schrittwieser and colleagues at DeepMind described MuZero, which planned using a model it learned rather than rules it was given.

  2. AlphaZero learns chess and shogi from the rules alone

    DeepMind posted a paper describing AlphaZero, one algorithm that, given only the rules, reached superhuman play at chess, shogi and Go within 24 hours of self-play training and beat the reigning computer champion program in each game.

  3. AlphaGo Zero learns Go from self-play alone

    DeepMind announced AlphaGo Zero, trained only by playing itself from random play, with no human games and no hand-crafted features. After three days it beat the version that had defeated Lee Sedol, 100 games to nil. Nature published the work the next day.

  4. Nature publishes the paper behind AlphaGo

    Nature published “Mastering the game of Go with deep neural networks and tree search” by David Silver, Demis Hassabis and 18 co-authors at Google DeepMind. It disclosed that AlphaGo had already beaten the European champion Fan Hui five games to nil in October 2015.