Register and share your invite link to earn from video plays and referrals.

Shaheer
@shaheersan
1.4K Following    2.7K Followers
Legendary discussions at the office. @rronak_ discussed applying and scaling up SDPO for continual learning in the real world to prevent context degradation and more @kevingu gave a talk on learning from production traces for knowledge-work tasks that have no direct verifier: how to generate reward signal without ground truth and keep organizational context current as the work changes @matthewjsargent explained how to stabilize asymmetric self play and new research fields in open ended learning Thanks to all those who came out and asked great questions we’ll be hosting more of these soon!
Show more