LLMs with scaffolds have lagged on text-to-SQL, a task that relies on human judgment. By folding expert judgment into every part of RLVR on Tinker,
@maxYuxuanZhu and
@ddkang (UIUC and Bridgwater) trained the first text-to-SQL model to beat the human mark.