登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

λux
@novasarc01
tensor shepherd in a non-euclidean pasture | grazing on cuda cores
参加 July 2024
3.5K フォロー中    22.5K ファン
some cool things i liked about verifiers v1: - my favorite shift from v0 to v1 is the move from an environment-centric abstraction to a rollout-centric one. v1 has this clean abstraction: taskset × harness × runtime → trace tasksets owning both data and scoring is much cleaner. - making the trace a first-class artifact is the biggest upgrade. also liked how graph-based trace storage avoids duplicating shared prefixes and makes long-horizon trace analysis much more practical. - being able to run the same taskset under different harnesses (kimi-code, rlm, codex, etc.) makes evaluations much more useful.
もっと見る