가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Adithya S K
@adithya_s_k
Scaling RL Envs @huggingface 🤗 • Founded @cognitivelab_ai Prev : Research @MSFTResearch • ML @apple • 22
가입 June 2020
2.4K 팔로잉 중    15.3K 팬
Very few people realize how closely RL environments and agent simulations go hand in hand. Most RL environments today are still static snapshots. Take coding or tool use: the repo, tools, APIs, data, etc. are already set up, and the agent acts on that existing state. But as environments get more sophisticated, you need to simulate the world around the agent: code reviews, email/message responses, user feedback, other agents, changing state... Doing this realistically, while keeping the environment verifiable and not reward-hackable, is a really fun problem to solve.
더 보기