註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Shaheer
@shaheersan
加入 July 2023
1.4K 正在關注    2.7K 粉絲
Legendary discussions at the office. @rronak_ discussed applying and scaling up SDPO for continual learning in the real world to prevent context degradation and more @kevingu gave a talk on learning from production traces for knowledge-work tasks that have no direct verifier: how to generate reward signal without ground truth and keep organizational context current as the work changes @matthewjsargent explained how to stabilize asymmetric self play and new research fields in open ended learning Thanks to all those who came out and asked great questions we’ll be hosting more of these soon!
顯示更多