注册并分享邀请链接,可获得视频播放与邀请奖励。

Azalia Mirhoseini
@Azaliamirh
Founder @RicursiveAI, Asst. Prof. of CS at Stanford. Prev: DeepMind, Anthropic, Brain. Co-Creator of MoEs, AlphaChip, Test Time Scaling.
加入 May 2013
626 正在关注    20.3K 粉丝
Check out TRACE, a new self-improvement approach where the agent identifies the missing capabilities behind its own failures and trains itself to address them. TRACE-trained Qwen3.6-27B reaches 73.2% on SWE-bench Verified, outperforming much larger models like Codex 5.2 and GLM 5, while beating GRPO and GEPA with <1/4 the training rollouts. By contrasting successful and failed trajectories, TRACE identifies its own weaknesses (such as bug localization or retrieval of the correct doc) and creates new synthetic environments to fix them. The result is a transferable and sample-efficient synthetic env / data generation + fine-tuning pipeline for agentic tasks. Great work led by @TarunSures41845 and @hangoo_kang!
显示更多
“TRACE: Capability-Targeted Agentic Training” got Spotlight @ ICML AIWILD 🎉 Beats direct RL, GEPA, & synthetic-agent data on SWE-Bench Verified and τ²-Bench. TRACE-Qwen3.6-27B tops GPT-5.2-Codex, GLM 5, & Claude 4.5 Sonnet on SWE-Bench. Co-led with @TarunSures41845. Thanks to @JonSaadFalcon and our advisor @Azaliamirh. Details below 👇
显示更多
0
13
501
53
转发到社区