登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Adithya S K
@adithya_s_k
Scaling RL Envs @huggingface 🤗 • Founded @cognitivelab_ai Prev : Research @MSFTResearch • ML @apple • 22
参加 June 2020
2.4K フォロー中    15.4K ファン
Very few people realize how closely RL environments and agent simulations go hand in hand. Most RL environments today are still static snapshots. Take coding or tool use: the repo, tools, APIs, data, etc. are already set up, and the agent acts on that existing state. But as environments get more sophisticated, you need to simulate the world around the agent: code reviews, email/message responses, user feedback, other agents, changing state... Doing this realistically, while keeping the environment verifiable and not reward-hackable, is a really fun problem to solve.
もっと見る