登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

(((ل()(ل() 'yoav))))👾
@yoavgo
参加 May 2009
2.2K フォロー中    89.6K ファン
"so harness capabilities are increasingly shifting into the model itself." --> many tnings one does with a harness can be baked in to a model by running the model with the harness' hints and behaviors, and then training it as if they weren't there. like "harness distillation"
もっと見る
GPT-6 Astra represents a step-function change in model capability for interactive reasoning problems. It scores 66% on ARC-AGI-3 using our standard harness, and nearly 100% with a continuous conversation harness and custom compaction, at a cost of roughly $360 per game. In fact, the continuous harness version significantly outperforms our human baseline in action efficiency across almost all levels. When we examined the reasoning chains to understand how the model operates, we found it performing highly efficient, on-the-fly symbolic world modeling for each game and level. It goes as far as developing its own shorthand DSL to represent in-game situations -- essentially a game-specific algebraic notation. Overall, Astra exhibits symbolic modeling behaviors we had previously only seen with sophisticated harnesses -- so harness capabilities are increasingly shifting into the model itself. We see Astra as a major breakthrough in model intelligence. Read our post on Astra and what these results mean:
もっと見る