注册并分享邀请链接,可获得视频播放与邀请奖励。

Xiuyu Li
@sheriyuo
Researcher @StepFun_ai | Working on long-horizon tasks | Prev @RUC1937 | Opinions are my own
加入 February 2026
1.8K 正在关注    13.4K 粉丝
The paper argues that for any ensemble whose final output must be one of the member models' answers, including routing, voting, and MoA, the accuracy is fundamentally capped by the co-failure rate β: acc ≤ 1 − β. It also shows that the commonly reported mean pairwise error correlation (ρ) is insufficient to characterize β, so low correlation alone does not imply large ensemble gains. When Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents
显示更多