Back
Recorded games become decision examples: the acting player’s view, action taken and game result. Imitation learning trains the base model’s policy and value; the daily games also refresh archetype decklist data.