가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Amanda Huang
@amandaH_333
LLM MLE @MiniMax_AI(ex-quant @RBC Capital Market, ex-Tiktoker) Role-play&Character training &Agent Climber-free solo someday not your typical algo guy
가입 October 2023
135 팔로잉 중    254 팬
Real self-correction isn’t asking a model to “think again.” It’s building a generate–verify–repair loop grounded in independent evidence, structured feedback, and hard stop conditions. Its effectiveness depends far more on the quality of the verifier and external feedback than on swapping models or adding more agent roles
더 보기