注册并分享邀请链接,可获得视频播放与邀请奖励。

Yifan Wu
@yifannnwu
吴奕凡; AI Research Scientist @Meta | Ph.D. @penn @picslupenn @GRASPlab.
加入 November 2016
490 正在关注    1.4K 粉丝
RL training for long-horizon tasks is still mysterious, and we took a baby step forward! 🧑‍🍳
Excited to share that SandMLE has been accepted by #COLM2026#! We introduced a multi-agent framework that generates diverse synthetic MLE environments to enable the large-scale on-policy RL training. See you then in SF!
显示更多