SW StudyWalks

Artificial Intelligence  /  AI 0158  ·  Atom · ~20 seconds

AlphaGo: Judgment Grown

Video not yet published
to the StudyWalks catalog
State

AlphaGo beat the world Go champion by combining tree search with learned judgment: networks trained on human games, then on play against itself.

Show

By its maker's account: a policy network proposing moves, a value network judging positions, sharpened by self-play.

Watch for

Where a written judge can be read and argued with, a grown judge cannot — a cost with its own unit later.