가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

DailyPapers
@HuggingPapers
Tweeting interesting papers submitted at Submit your own at and link models/datasets/demos to it!
가입 March 2025
4 팔로잉 중    20.1K
V-Zero: answer-label-free visual reasoning It uses on-policy distillation with contrastive evidence gating. It trains 5x faster than SFT and 10x faster than RL. 4B model is on Hugging Face.
더 보기