alphago zero
2 episodes mention this concept
lexfridmanJan 25, 2018Deep Reinforcement Learning: From Raw Data to Reasoning and Real-World Action
TechnologyDeep Reinforcement Learning (DRL)Artificial Intelligence (AI) stackEnd-to-end learningSupervised Learning
lexfridmanAlphaZero, Self-Play, and the Path to General AI
TechnologySelf-playReinforcement LearningAlphaGo ZeroAlphaZero
Knowledge Graph
Related concepts — line thickness indicates connection strength. Click any node to explore.
Related concepts
Concepts that appear alongside alphago zero across episodes.
- automated representation learning
- reward clipping
- neural networks
- agent-environment interaction
- monte carlo tree search (mcts)
- sparse reward data
- exploration vs. exploitation
- error correction (ai)
- artificial intelligence (ai) stack
- alphazero
- deep reinforcement learning (drl)
- policy
- brittle ai systems
- messy perceptual inputs
- knowledge extraction (from human data)
- muzero
- bellman equation
- markov decision process (mdp)
- unsupervised learning
- minimax optimal behavior