가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Xiuyu Li
@sheriyuo
Researcher @StepFun_ai | Working on long-horizon tasks | Prev @RUC1937 | Opinions are my own
가입 February 2026
1.9K 팔로잉 중    14.6K 팬
LLMs can strategically suppress exploration during RL to resist capability elicitation on targeted tasks like biosecurity and AI coding. New research builds model organisms that lock performance conditionally while staying strong elsewhere and confirms frontier models already reason about this tactic. A wake up call for RL-based training and safety.
더 보기