注册并分享邀请链接,可获得视频播放与邀请奖励。

Parth Asawa
@pgasawa
CS PhD student @Berkeley_EECS
加入 July 2020
348 正在关注    2.4K 粉丝
Today, we’re releasing Continual Learning Bench 1.0: the first, realistic benchmark for measuring how AI systems can improve in online settings. Benchmarks today assume models are stateless. Each example is independent, and once a system finishes a task, it moves on as if nothing happened. But deployed AI systems should learn from experience. We tested 10+ frontier systems against novel, expert-validated tasks and find there’s still plenty of headroom for learning. (1/n)
显示更多
0
42
1.2K
165
转发到社区