註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Brendan (can/do)
@BrendanFoody
ceo @mercor | organizing human intelligence
加入 March 2020
542 正在關注    28.7K 粉絲
At this important juncture, building better evaluations across a diverse set of domains is the most impactful way to align models. We need subject matter experts across cyber, biology, chemistry, and many other industries to build benchmarks that help pace the frontier.
顯示更多
Mercor is committing $5M to a new AI Safety Fund. One of the biggest challenges the industry faces today is addressing whether frontier AI is safe enough to deploy. This is why we believe investing today in safety research, evals, and verification is critical. We want to work with AI researchers who are focused on critical safety risks, including: - Misalignment: deceptive alignment, reward hacking, scheming - Sandbox escape and agent containment failures - Evaluation awareness: models that behave differently when tested - Interpretability and scalable oversight - Red-teaming methodology and safety eval design We will fund the researcher hours, API credits, and travel. We will also cover the cost to work with our network of 5M+ experts red-teaming, grading and annotation, and free use of our evals and analysis platform. Grants are open to independent researchers, non-profits, and academics.
顯示更多
0
10
77
10
轉發到社區