Many mathematicians are complaining that AI proofs are incomprehensible and therefore does not further our understanding of mathematics.
To me this seems like a relatively tractable problem for the AI labs to solve, no? Like couldn't you have some sort of reward model/judge that measures how easy to understand a proof is to understand and use that to post-train or guide the models?
Perhaps I am oversimplifying this... if so, please explain why this would be hard?