Most upvoted papers on
@huggingface this week
Harness Handbook - Making evolving agent harnesses readable and editable
LongStraw - Long-context RL beyond 2M tokens on fixed GPU budget
Weak-to-Strong Generalization via Direct On-Policy Distillation
Boogu-Image-0.1 - Open-source unified multimodal understanding and generation
VideoChat3 - Fully open video MLLM for efficient generalist understanding
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agents
ABot-N1 - Toward a general visual language navigation foundation model
Find them below: