AlphaZero

E40166

AlphaZero is a DeepMind-developed artificial intelligence system that mastered complex games like chess, shogi, and Go through self-play reinforcement learning without human-crafted strategies.

AI illustration

How this image was made

AI-generated illustration of AlphaZero

This AI-generated illustration was produced by black-forest-labs/FLUX.2-dev (1024x1024) from a prompt written by openai/gpt-oss-120b from the entity's label + description.

Prompt

Generate an image of AlphaZero (AlphaZero is a DeepMind-developed artificial intelligence system that mastered complex games like chess, shogi, and Go through self-play reinforcement learning without human-crafted strategies.)

All labels observed (6)

How this entity was disambiguated

Statements (53)

Predicate Object
instanceOf artificial intelligence system ⓘ
game‑playing program ⓘ
architectureType deep neural network with Monte Carlo tree search ⓘ
basedOn Monte Carlo tree search ⓘ
deep learning ⓘ
reinforcement learning ⓘ
contrastWith programs relying on human expert knowledge ⓘ
traditional chess engines using alpha‑beta search ⓘ
countryOfOrigin United Kingdom ⓘ
creatorOrganizationType AI research lab ⓘ
defeated Elmo shogi engine ⓘ
Stockfish 8 ⓘ
linked to: Stockfish

previous Go programs based on AlphaGo Zero ⓘ
designedFor Go ⓘ
chess ⓘ
shogi ⓘ
developer DeepMind ⓘ
Google DeepMind ⓘ
linked to: DeepMind
doesNotUse endgame tablebases for search guidance ⓘ
human‑crafted opening books ⓘ
evaluationFunction learned value function ⓘ
field artificial intelligence ⓘ
computer Go ⓘ
computer chess ⓘ
computer shogi ⓘ
machine learning ⓘ
firstPublicAnnouncementDate 2017-12-06 ⓘ
firstPublicAnnouncementYear 2017 ⓘ
gameRepresentation board positions encoded for neural networks ⓘ
generalizationProperty single algorithm applied to multiple games ⓘ
hardwareUsed TPUs ⓘ
learningObjective maximize expected game outcome ⓘ
learningParadigm tabula rasa learning ⓘ
notableFor mastering Go through self‑play ⓘ
mastering chess through self‑play ⓘ
mastering shogi through self‑play ⓘ
outperforms AlphaGo Zero ⓘ
Elmo ⓘ
Stockfish ⓘ
parentProject AlphaGo project ⓘ
linked to: AlphaGo
policyRepresentation probability distribution over moves ⓘ
publicationTitle A general reinforcement learning algorithm that masters chess, shogi, and Go through self‑play ⓘ
linked to: AlphaZero
publishedIn Science ⓘ
rewardSignal game result win‑draw‑loss ⓘ
searchGuidance policy network priors ⓘ
value network evaluations ⓘ
searchTechnique Monte Carlo tree search guided by neural networks ⓘ
trainingDataSource self‑generated game data ⓘ
trainingMethod self‑play ⓘ
trainingRegime self‑play reinforcement learning without human examples ⓘ
uses neural networks ⓘ
policy network ⓘ
value network ⓘ

How these facts were elicited

Referenced by (24)

Full triples — surface form annotated when it differs from this entity's canonical label.

DeepMind → knownFor → AlphaZero ⓘ
DeepMind → developed → AlphaGo Zero ⓘ
linked to: AlphaZero
DeepMind → developed → AlphaZero ⓘ
Demis Hassabis → knownFor → AlphaZero ⓘ
AlphaGo → successor → AlphaGo Zero ⓘ
linked to: AlphaZero
AlphaGo → successor → AlphaZero ⓘ
AlphaGo → inspired → AlphaGo Zero ⓘ
linked to: AlphaZero
AlphaGo → inspired → AlphaZero ⓘ
David Silver → knownFor → AlphaZero ⓘ
David Silver → notableWork → AlphaZero ⓘ
David Silver → notablePaper → Mastering chess and shogi by self-play with a general reinforcement learning algorithm ⓘ
linked to: AlphaZero
AlphaStar → relatedTo → AlphaZero ⓘ
AlphaZero → publicationTitle → A general reinforcement learning algorithm that masters chess, shogi, and Go through self‑play ⓘ
linked to: AlphaZero
MuZero → inspiredBy → AlphaZero ⓘ
MuZero → comparedTo → AlphaZero ⓘ
Demis Hassabis → knownFor → AlphaZero ⓘ
subject linked to: Demis
Ioannis Antonoglou → knownFor → AlphaZero ⓘ
Elmo shogi engine → usedBy → DeepMind AlphaZero ⓘ
linked to: AlphaZero
Elmo shogi engine → comparedWith → AlphaZero (shogi) in DeepMind experiments ⓘ
linked to: AlphaZero
Matthew Sadler → hasWrittenAbout → AlphaZero ⓘ
Leela Chess Zero → inspiredBy → AlphaZero ⓘ
Thore Graepel → notableWork → AlphaZero ⓘ