註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Joe Hansen
@joehansen
Tesla | SpaceX | Starlink | xAI | Grok | Terafab Thoughts & Analysis
加入 July 2014
4.3K 正在關注    22.6K 粉絲
Grok 4.6 just took #1# on CursorBench. Not only the highest score. It did it at a fraction of the cost of the models sitting right behind it. That combination is the real signal. Top-tier results are one thing. Top-tier results that stay cheap enough to run for long agentic coding sessions are something else. I see this as the practical edge that matters for real work. Benchmarks are useful. Sustained performance at low cost is what actually gets used.
顯示更多
0
42
500
79
轉發到社區