MuZero masters games without being told the rules
The system learned to predict only value, policy and reward rather than the environment's rules, then matched AlphaZero at Go, chess and shogi and set a new Atari benchmark.
Google DeepMindModels & capabilities