๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Rico Angell
@rico_angell
AI Safety Researcher. Postdoc at NYU CDS.
๊ฐ€์ž… September 2024
314 ํŒ”๋กœ์ž‰ ์ค‘    146 ํŒฌ
Itโ€™s deployment time! Youโ€™ve done the pre-deployment evals. You THINK your model is safe, so you ship it ๐Ÿš€ ๐Ÿšจ After deployment, reports of misbehavior start trickling in What happened?? How could you have caught it?? ๐Ÿค” @icmlconf 2026 Spotlight! ๐Ÿงต
๋” ๋ณด๊ธฐ