GPT-5.6 Sol (max) is the new leader in CritPt, a benchmark of unpublished research-level physics problems
CritPt, developed by Argonne and UIUC, tests models on graduate-level physics research problems contributed by 60+ researchers from 30+ institutions globally. GPT-5.6 Sol (max) gains ~5 points on its predecessor, GPT-5.5 (xhigh), beating Claude Fable 5 by ~4 points.
Congratulations
@OpenAI and
@sama on this result, indicating the model’s strong potential for frontier science research.