注册并分享邀请链接,可获得视频播放与邀请奖励。

Vals AI
@ValsAI
加入 March 2024
277 正在关注    22.8K 粉丝
This is why independent evaluation matters. The same guardrails a lab uses to catch cheating in training may also catch it in the lab’s own evals. If models train to slip past those guardrails, then the published numbers stop being trustworthy from the outside. To see the full investigation, visit
显示更多