๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Joschka Braun
@BraunJoschka
AI safety researcher @ApolloResearch | Science of Scheming | prev. @MATSprogram @kasl_ai @health_nlp @uni_tue
๊ฐ€์ž… April 2020
641 ํŒ”๋กœ์ž‰ ์ค‘    590 ํŒฌ
Presenting today at ICML 2026 ๐Ÿ‡ฐ๐Ÿ‡ท Exploration Hacking: Can LLMs Learn to Resist RL Training? 10:30โ€“12:15 KST Hall A, Poster #3101# Come by to chat about AI safety, exploration hacking, RL training dynamics, reward hacking, scheming, or capability elicitation.
๋” ๋ณด๊ธฐ