Replying to
@bgavran@mathstodon.xyz “The system AlphaZero utilises the game structure during training and reaches superhuman ELO (>3400) with ~30x fewer parameters than GPT-4 (<60 million vs 1.8 trillion).” — I guess that’s supposed to be a “30.000x fewer”? (Which is quite immense.)