가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Goodfire
@GoodfireAI
Using interpretability to understand, learn from, and design AI.
가입 August 2024
31 팔로잉 중    26.3K
Our team spent months developing RLFR, our method which uses probes on a model's internals as reward signals for RL. Silico reproduced it in 2 days, reducing hallucinations in Qwen3-8B by 37% without capability loss. (3/6)
더 보기