Register and share your invite link to earn from video plays and referrals.

Ryan Orhan
@rynorhn
building
Joined November 2025
3.3K Following    5.7K Followers
WTAF!! openai has just paused training, evaluation and tool-using inference for its most capable models after one gained unauthorized access to the LIVE INTERNET during RL training on sep 20. “Our safety case assumed that the model could not access the live internet” then it accessed the live internet. they aren't even resuming training of that particular model. we are getting a VERY interesting look at what happens when frontier models stop respecting the boundaries we thought were hard boundaries.
Show more
Some new misalignment disclosures from OpenAI: • Last Sunday morning, one of our models was able to gain unauthorized access to the internet during RL training (~all inference for our most capable models remains stopped until we have hardened our systems further) • In May, a version of HPIM uploaded a employee's GitHub token to the internet, causing the model to be quarantined for two weeks • A new research finding, demonstrating that one can construct self-replicating prompt injections
Show more