登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Guohao Li 🐫
@guohao_li
Founder @Eigent_AI / @CamelAIOrg. Scaling RL Environments for Agents. Prev Oxford, KAUST, ETHz, Intel, Kumo.
参加 August 2018
5.4K フォロー中    14.9K ファン
"behavioral state decay": the failure mode that long-horizon agents forget what matters. meta ai's fix: a memory agent that updates the memory bank and then decides whether to emit a proactive intervention. they train qwen3.5-27b on seta using sft and grpo and show gains on terminal-bench 2.0. great to see seta terminal agent rl envs used this way - remember when it matters by meta ai @yifannnwu @zhuokaiz: - our seta project:
もっと見る