注册并分享邀请链接,可获得视频播放与邀请奖励。

Amanda Huang
@amandaH_333
LLM MLE @MiniMax_AI(ex-quant @RBC Capital Market, ex-Tiktoker) Role-play&Character training &Agent Climber-free solo someday not your typical algo guy
加入 October 2023
135 正在关注    254 粉丝
Real self-correction isn’t asking a model to “think again.” It’s building a generate–verify–repair loop grounded in independent evidence, structured feedback, and hard stop conditions. Its effectiveness depends far more on the quality of the verifier and external feedback than on swapping models or adding more agent roles
显示更多