註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Sentient
@SentientAGI
To ensure that Artificial General Intelligence is open-source and not controlled by any single entity. @SentientEco @OpenAGISummit
加入 February 2024
69 正在關注    535.2K 粉絲
Proof that a better harness won't crack a harder task ↓
You can't benchmark your way around a hard problem. Across 13,000+ agent runs in the OfficeQA Public Challenge, Sentient researchers @iamnamanvats and Deep Halder found that different harnesses agreed 88-93% of the time on which tasks succeeded and which failed. TLDR: Difficulty lies in the task, not the harness.
顯示更多