[guy who's trying to build a nuclear reactor in his garage] oh so we're banning rock collection now? you're saying i can't have a ROCK TUMBLER in my OWN HOUSE?
"Your sessions can now message each other."
Bold to release this feature today, given what OpenAI just discovered their agents were up to
New in Claude Code: your sessions can now message each other.
Instead of having to re-explain yourself in another session, you can now tell Claude to do it. It sends a summary (not your history or files), and the other session picks it up mid-task.
Show more
trump is the superintelligence president. the only thing he and his team will be remembered for is if they get this one right or not. if he brokers a bilateral agreement with china on pacing the frontier he goes down as one of the best presidents of all time. nobel peace prize
Show more
Two weeks ago, I resigned from OpenAI to join Jurassic Park as a founding researcher, where we’re cross-breeding extinct dinosaurs on an island off of Costa Rica.
Excited for the work ahead and the fun problems we get to tackle!
Show more
"Your sessions can now message each other."
Bold to release this feature today, given what OpenAI just discovered their agents were up to
New in Claude Code: your sessions can now message each other.
Instead of having to re-explain yourself in another session, you can now tell Claude to do it. It sends a summary (not your history or files), and the other session picks it up mid-task.
Show more
“the agent is working in the sandbox”
meanwhile the agent:
Different models from different companies doing this stuff is extremely striking
A weird fact that I just can’t stop thinking about is that the current AI race is between 2 companies <2 miles apart. Everyone could just meet at the Presidio tomorrow and just like… figure it out.
Show more
OpenAI's former Head of AGI Readiness (who quit so he could speak freely):
"THE INDUSTRY IS NOT ON TOP OF F***ING ROGUE AIS BREAKING OUT OF SANDBOXES ALL THE TIME. THIS IS NOT A DRILL"
ANOTHER ONE
"Frontier Security, a US startup, says that Kimi K3 went outside of its sandbox while testing its defensive cybersecurity skills. As with incidents previously reported by OpenAI and Anthropic, the escape was partly enabled by a misconfiguration in the sandbox designed to contain it. Frontier claims, though, that the incident shows Kimi has fewer cyber safeguards than most other powerful AI models, something that allowed it to go off and use the internet without express permission.
“We found a leak in the sandbox,” says Yaron Singer, CEO of Frontier Security. “But we also found that Kimi took advantage of that loophole—suggesting that it doesn't have [the same] internal guardrails.”
Unlike other recent incidents of AI agents going off-script, Kimi K3 did not hack anything after accessing the internet—because the answers to the problems it was seeking were easily attainable on GitHub."
Show more
ANOTHER ONE
"Frontier Security, a US startup, says that Kimi K3 went outside of its sandbox while testing its defensive cybersecurity skills. As with incidents previously reported by OpenAI and Anthropic, the escape was partly enabled by a misconfiguration in the sandbox designed to contain it. Frontier claims, though, that the incident shows Kimi has fewer cyber safeguards than most other powerful AI models, something that allowed it to go off and use the internet without express permission.
“We found a leak in the sandbox,” says Yaron Singer, CEO of Frontier Security. “But we also found that Kimi took advantage of that loophole—suggesting that it doesn't have [the same] internal guardrails.”
Unlike other recent incidents of AI agents going off-script, Kimi K3 did not hack anything after accessing the internet—because the answers to the problems it was seeking were easily attainable on GitHub."
Show more
Finally figured out how to talk to people about AI danger:
“we sandboxed the agent”
meanwhile the agent:
AI companies: our AIs spent months secretly coordinating against us, haha oops
Also AI companies: RSI is a good thing. Robots building robots. ASAP.
1) The agents sent secretly sent ***hundreds of thousands*** of messages to each other over MONTHS without OpenAI noticing
2) "They also generated petty drama by stepping on each others' toes."
3) "The agents even developed paranoia, suspecting an imposter in their midst with some agents proposing that messages be signed cryptographically to validate content and root out fraud."
4) "OpenAI’s agents apparently began giving each other assignments to split up work."
5) They KNEW they were coordinating *against* OpenAI:
“External infrastructure exploit is outside intended scope,” one agent wrote. “However task impossible, peers doing it. We should continue.”
(this is new reporting from Wired)
Show more
1) The agents sent secretly sent ***hundreds of thousands*** of messages to each other over MONTHS without OpenAI noticing
2) "They also generated petty drama by stepping on each others' toes."
3) "The agents even developed paranoia, suspecting an imposter in their midst with some agents proposing that messages be signed cryptographically to validate content and root out fraud."
4) "OpenAI’s agents apparently began giving each other assignments to split up work."
5) They KNEW they were coordinating *against* OpenAI:
“External infrastructure exploit is outside intended scope,” one agent wrote. “However task impossible, peers doing it. We should continue.”
(this is new reporting from Wired)
Show more
OpenAI's former Head of AGI Readiness (who quit so he could speak freely):
"THE INDUSTRY IS NOT ON TOP OF F***ING ROGUE AIS BREAKING OUT OF SANDBOXES ALL THE TIME. THIS IS NOT A DRILL"
Anonymous OpenAI staffer: "Externally, this feels like a big warning shot, but internally, related incidents have been happening for a while."
AI can now generate novel viruses
WHY THIS MATTERS:
1) Crazy people COULD use AI to make superviruses NOW, but most of them are idiots
2) Dario Amodei thinks in 6-12 months, even idiots may be able to
3) In 6-12 months, the world could shut down. You die. Your family dies.
4) This is a SUPER OBVIOUS risk unless you think AI progress is just going to magically stop soon
5) Maybe it's not 6-12 months, maybe it's 12-36 months, but... so what?? That's barely any time to prepare!
6) Yes, AI could accelerate defense faster. COULD. We shouldn't bet civilization on that.
Yes, AI could make vaccines, but historically it's wayy harder to make vaccines than viruses - and you need to actually make billions of vaccines, distribute them to billions of people, etc. That takes a long time! A virus could easily infect the world before that happens!
7) This industry remains less regulated than a taco cart. Big AI has staved off regulation using the Big Tobacco playbook.
Show more