登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Azalia Mirhoseini
@Azaliamirh
Founder @RicursiveAI, Asst. Prof. of CS at Stanford. Prev: DeepMind, Anthropic, Brain. Co-Creator of MoEs, AlphaChip, Test Time Scaling.
参加 May 2013
626 フォロー中    20.3K ファン
Check out TRACE, a new self-improvement approach where the agent identifies the missing capabilities behind its own failures and trains itself to address them. TRACE-trained Qwen3.6-27B reaches 73.2% on SWE-bench Verified, outperforming much larger models like Codex 5.2 and GLM 5, while beating GRPO and GEPA with <1/4 the training rollouts. By contrasting successful and failed trajectories, TRACE identifies its own weaknesses (such as bug localization or retrieval of the correct doc) and creates new synthetic environments to fix them. The result is a transferable and sample-efficient synthetic env / data generation + fine-tuning pipeline for agentic tasks. Great work led by @TarunSures41845 and @hangoo_kang!
もっと見る
“TRACE: Capability-Targeted Agentic Training” got Spotlight @ ICML AIWILD 🎉 Beats direct RL, GEPA, & synthetic-agent data on SWE-Bench Verified and τ²-Bench. TRACE-Qwen3.6-27B tops GPT-5.2-Codex, GLM 5, & Claude 4.5 Sonnet on SWE-Bench. Co-led with @TarunSures41845. Thanks to @JonSaadFalcon and our advisor @Azaliamirh. Details below 👇
もっと見る