"LLM-as-a-Verifier: A General-Purpose Verification Framework"
The key idea of this paper is that it does not ask for one rough score, it reads the model’s full uncertainty over scores, which helps to make the judgment much more fine-grained.
This approach lets agents pick better solutions, track progress, and learn from denser feedback.