A Go-playing AI built with deep learning and reinforcement learning.
AlphaGo turns game night into boot camp. It plays itself all night, then makes Go champions sweat.
It used Deep RL to master Go. It made RL famous outside labs. It showed AI could plan tough moves.
RL
AlphaGo used repeated games and feedback to improve its strategy.
MCTS
AlphaGo used MCTS to judge positions and choose stronger moves.
Deep RL
AlphaGo became the breakout example of Deep RL.
AlphaZero
AlphaZero built on AlphaGo and spread the idea to more board games.