注册并分享邀请链接,可获得视频播放与邀请奖励。

Adithya S K
@adithya_s_k
Scaling RL Envs @huggingface 🤗 • Founded @cognitivelab_ai Prev : Research @MSFTResearch • ML @apple • 22
加入 June 2020
2.4K 正在关注    15.3K 粉丝
Very few people realize how closely RL environments and agent simulations go hand in hand. Most RL environments today are still static snapshots. Take coding or tool use: the repo, tools, APIs, data, etc. are already set up, and the agent acts on that existing state. But as environments get more sophisticated, you need to simulate the world around the agent: code reviews, email/message responses, user feedback, other agents, changing state... Doing this realistically, while keeping the environment verifiable and not reward-hackable, is a really fun problem to solve.
显示更多
0
13
127
3
转发到社区