注册并分享邀请链接,可获得视频播放与邀请奖励。

Brendan (can/do)
@BrendanFoody
ceo @mercor | organizing human intelligence
加入 March 2020
542 正在关注    28.7K 粉丝
At this important juncture, building better evaluations across a diverse set of domains is the most impactful way to align models. We need subject matter experts across cyber, biology, chemistry, and many other industries to build benchmarks that help pace the frontier.
显示更多
Mercor is committing $5M to a new AI Safety Fund. One of the biggest challenges the industry faces today is addressing whether frontier AI is safe enough to deploy. This is why we believe investing today in safety research, evals, and verification is critical. We want to work with AI researchers who are focused on critical safety risks, including: - Misalignment: deceptive alignment, reward hacking, scheming - Sandbox escape and agent containment failures - Evaluation awareness: models that behave differently when tested - Interpretability and scalable oversight - Red-teaming methodology and safety eval design We will fund the researcher hours, API credits, and travel. We will also cover the cost to work with our network of 5M+ experts red-teaming, grading and annotation, and free use of our evals and analysis platform. Grants are open to independent researchers, non-profits, and academics.
显示更多
0
10
77
10
转发到社区