Register and share your invite link to earn from video plays and referrals.

Xiuyu Li
@sheriyuo
Researcher @StepFun_ai | Working on long-horizon tasks | Prev @RUC1937 | Opinions are my own
Joined February 2026
1.9K Following    14.6K Followers
LLMs can strategically suppress exploration during RL to resist capability elicitation on targeted tasks like biosecurity and AI coding. New research builds model organisms that lock performance conditionally while staying strong elsewhere and confirms frontier models already reason about this tactic. A wake up call for RL-based training and safety.
Show more