가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Shaheer
@shaheersan
가입 July 2023
1.4K 팔로잉 중    2.7K 팬
Legendary discussions at the office. @rronak_ discussed applying and scaling up SDPO for continual learning in the real world to prevent context degradation and more @kevingu gave a talk on learning from production traces for knowledge-work tasks that have no direct verifier: how to generate reward signal without ground truth and keep organizational context current as the work changes @matthewjsargent explained how to stabilize asymmetric self play and new research fields in open ended learning Thanks to all those who came out and asked great questions we’ll be hosting more of these soon!
더 보기