登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Goodfire
@GoodfireAI
Using interpretability to understand, learn from, and design AI.
参加 August 2024
31 フォロー中    26.3K ファン
Our team spent months developing RLFR, our method which uses probes on a model's internals as reward signals for RL. Silico reproduced it in 2 days, reducing hallucinations in Qwen3-8B by 37% without capability loss. (3/6)
もっと見る