註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Goodfire
@GoodfireAI
Using interpretability to understand, learn from, and design AI.
加入 August 2024
31 正在關注    26.3K 粉絲
Our team spent months developing RLFR, our method which uses probes on a model's internals as reward signals for RL. Silico reproduced it in 2 days, reducing hallucinations in Qwen3-8B by 37% without capability loss. (3/6)
顯示更多
0
8
357
16
轉發到社區