註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Mercor
@mercor
Organizing human intelligence to power the AI economy.
加入 April 2021
31 正在關注    24.5K 粉絲
Data is the most important ingredient in post-training. The Mercor Research team focuses on making every hour of expert work yield the most model improvement, through better learning algorithms for knowledge work and automated, domain-specific post-training. We're also committed to doing open source research. That's why we're publishing a RL training guide for Qwen3.5-397B in collaboration with the SkyRL team. Find out how we raised Pass@1 on APEX-Agents from 16% to 27% and explore the full training script, model weights, and eval traces. Want to do this kind of work with us? We're hiring. Read the full blog post:
顯示更多