AlphaStar Deep Dive: The Nature Paper (Vinyals et al., 2019)
Question-by-question drill on the AlphaStar paper: the two-phase training pipeline, network architecture, RL update (TD(lambda), V-trace, UPGO, KL), league training with PFSP and exploiters, fairness constraints, and transfer to complex multi-agent systems.
- Status
- Not started
- Quiz recall
- 0%
- Duration
- ~45 min
- Remaining
- 22 questions