I see many people surprised that AI is solving most problems in Mathematics and coding (Navier-Stokes, Erdos, superhuman hacking skills). This ties into two topics 1. RSI and 2. Verifiability
The answer to this is that these problems are verifiable. AI will solve in my view any verifiable problem through RSI. Math and coding are fundamentally verifiable. We can prove a math solution is correct or a coding solution is correct (within reason here).
In RL, it is easy to train on verifiable tasks, and you can continue scaling this out. Generating more and more verifiable tasks. Now training on non verifiable tasks is much harder. How do you train a model to output "pretty" things? Pretty is subjective and not really verifiable like Math.