注册并分享邀请链接,可获得视频播放与邀请奖励。

Composio
@composio
Your agent is smart. Its tools should be too. Check
加入 October 2023
49 正在关注    24.6K 粉丝
We tested 6 AI models on 30 challenging agent tasks: GPT-6 Astra, Opus 5.5, GPT-6 Sol, Pareto 26.9, DeepSeek V4 Pro, and GLM 5.3 Flash. Sol matched Opus’s score, finished faster, and cost about a quarter as much per successful task. Here’s how all 6 models compared 🧵🧵🧵
显示更多
0
35
221
16
转发到社区