muzero
2 episodes mention this concept
lexfridmanApr 3, 2020David Silver on AlphaGo, AlphaZero, and the Reinforcement Learning Path to Artificial Intelligence
TechnologyReinforcement Learning (RL)Deep Reinforcement LearningAlphaGoAlphaZero
lexfridmanAlphaZero, Self-Play, and the Path to General AI
TechnologySelf-playReinforcement LearningAlphaGo ZeroAlphaZero
Knowledge Graph
Related concepts — line thickness indicates connection strength. Click any node to explore.
Related concepts
Concepts that appear alongside muzero across episodes.
- alphazero
- alphago zero
- knowledge acquisition bottleneck
- error correction (ai)
- reward signal
- policy (rl)
- alphastar
- game of go
- brittle ai systems
- messy perceptual inputs
- knowledge extraction (from human data)
- deep reinforcement learning
- minimax optimal behavior
- value function
- self-play
- heuristic search
- artificial intelligence (ai)
- reinforcement learning (rl)
- information processing systems
- model-based reinforcement learning