註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Jessy Lin
@realJessyLin
cofounder @EngramLab | prev PhD @Berkeley_AI
加入 March 2013
1K 正在關注    6.3K 粉絲
Building a good benchmark for continual learning takes a lot of thought -- it's non-trivial to make it hard in the ways that matter. excited to see people working towards this
Today, we’re releasing Continual Learning Bench 1.0: the first, realistic benchmark for measuring how AI systems can improve in online settings. Benchmarks today assume models are stateless. Each example is independent, and once a system finishes a task, it moves on as if nothing happened. But deployed AI systems should learn from experience. We tested 10+ frontier systems against novel, expert-validated tasks and find there’s still plenty of headroom for learning. (1/n)
顯示更多