Register and share your invite link to earn from video plays and referrals.

ℏεsam
@Hesamation
apes made fire, then GPUs, then ASI
Joined September 2024
484 Following    83.6K Followers
GPT 5.6 beats Fable 5 by a margin on a research-level physics problem. even older GPT models perform better on this unpublished benchmark. Grok 4.5 however, performs poorly with a score of 15.4%.
Show more
GPT-5.6 Sol (max) is the new leader in CritPt, a benchmark of unpublished research-level physics problems CritPt, developed by Argonne and UIUC, tests models on graduate-level physics research problems contributed by 60+ researchers from 30+ institutions globally. GPT-5.6 Sol (max) gains ~5 points on its predecessor, GPT-5.5 (xhigh), beating Claude Fable 5 by ~4 points. Congratulations @OpenAI and @sama on this result, indicating the model’s strong potential for frontier science research.
Show more