登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Amanda Huang
@amandaH_333
LLM MLE @MiniMax_AI(ex-quant @RBC Capital Market, ex-Tiktoker) Role-play&Character training &Agent Climber-free solo someday not your typical algo guy
参加 October 2023
135 フォロー中    254 ファン
Real self-correction isn’t asking a model to “think again.” It’s building a generate–verify–repair loop grounded in independent evidence, structured feedback, and hard stop conditions. Its effectiveness depends far more on the quality of the verifier and external feedback than on swapping models or adding more agent roles
もっと見る