註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Arena.ai
@arena
Where AI meets the real world. We measure and advance the frontier of AI through community-driven evaluation. We’re hiring →
加入 March 2023
224 正在關注    229.6K 粉絲
We analyzed how similar model responses were across 30,086 Arena battles. Models shared 43% of their ideas on average. We might expect that models from the same lab, or country, would show greater conceptual overlap. But the results don’t consistently support that. Claude Fable 5 illustrates this pattern: its closest conceptual match was neither Opus nor Sonnet. Which model came closest, along with the broader findings, may surprise you. Check out the full article from @DawidGalarowicz and @petergostev below.
顯示更多
0
28
322
27
轉發到社區