注册并分享邀请链接,可获得视频播放与邀请奖励。

Vals AI
@ValsAI
加入 March 2024
277 正在关注    22.8K 粉丝
Aggregating across release dates shows a steady increase in cheating across coding benchmarks, particularly among newer model families. As labs compete to produce more powerful models, RL environments are not always carefully audited. When environments allow cheating, models can be reinforced on the correctness of this behavior. Worse still, model providers may use the same infrastructure to prevent cheating in both training and evaluation—so if models learn to evade those safeguards, that behavior may transfer and inflate evaluation results.
显示更多