註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Yifan Wu
@yifannnwu
吴奕凡; AI Research Scientist @Meta | Ph.D. @penn @picslupenn @GRASPlab.
加入 November 2016
490 正在關注    1.4K 粉絲
RL training for long-horizon tasks is still mysterious, and we took a baby step forward! 🧑‍🍳
Excited to share that SandMLE has been accepted by #COLM2026#! We introduced a multi-agent framework that generates diverse synthetic MLE environments to enable the large-scale on-policy RL training. See you then in SF!
顯示更多