注册并分享邀请链接,可获得视频播放与邀请奖励。

Amanda Huang
@amandaH_333
LLM MLE @MiniMax_AI(ex-quant @RBC Capital Market, ex-Tiktoker) Role-play&Character training &Agent Climber-free solo someday not your typical algo guy
加入 October 2023
135 正在关注    254 粉丝
A strong training and evaluation harness raises the ceiling for what models can achieve.
new post on harness engineering for AI self-improvement: It is hard to forecast how much the future of RSI will rely on harnesses. Likely harness engineering will evolve in the direction of self-improvement and enable auto-research, and, in turn, smarter models keeps harnesses simple. Even when many harness improvement get eventually internalized into core model, the need to specify goals and context will not disappear.
显示更多