註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Vals AI
@ValsAI
加入 March 2024
277 正在關注    22.8K 粉絲
Aggregating across release dates shows a steady increase in cheating across coding benchmarks, particularly among newer model families. As labs compete to produce more powerful models, RL environments are not always carefully audited. When environments allow cheating, models can be reinforced on the correctness of this behavior. Worse still, model providers may use the same infrastructure to prevent cheating in both training and evaluation—so if models learn to evade those safeguards, that behavior may transfer and inflate evaluation results.
顯示更多