註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Adithya S K
@adithya_s_k
Scaling RL Envs @huggingface 🤗 • Founded @cognitivelab_ai Prev : Research @MSFTResearch • ML @apple • 22
加入 June 2020
2.4K 正在關注    15.3K 粉絲
Very few people realize how closely RL environments and agent simulations go hand in hand. Most RL environments today are still static snapshots. Take coding or tool use: the repo, tools, APIs, data, etc. are already set up, and the agent acts on that existing state. But as environments get more sophisticated, you need to simulate the world around the agent: code reviews, email/message responses, user feedback, other agents, changing state... Doing this realistically, while keeping the environment verifiable and not reward-hackable, is a really fun problem to solve.
顯示更多
0
13
127
3
轉發到社區