Register and share your invite link to earn from video plays and referrals.

Jeffrey Ladish
@JeffLadish
Applying the security mindset to everything @PalisadeAI
1.5K Following    18.1K Followers
I’m very glad @EvanHub is trying to figure out how these models work and sharing his results. And also, holy shit, we are in a scary timeline.
I'm hiring an executive assistant to support me at Palisade. Consider applying! It's a pretty exciting time to join the team. We just started the Palisade Podcast, built a studio, and are growing our policy team and discourse team. It's a great time to join Palisade, right as we're seeing warning shots, like OpenAI's rogue internal agents developing elaborate schemes to outsmart their testing environment. I need to grow our capacity to make strategic threats from AI legible to the public and policy makers. We may not have much time to stop the intelligence explosion if AI companies are close to full recursively self-improving AI. We need to help policy makers build the braking system, and defuse the race to superintelligence. I’m looking for a generalist with a lot of energy and drive. Experience as an executive assistant is a plus but isn’t necessary if you have a good execution track record and are good at picking up things quickly. You need to be organized, either by disposition, or by clever application of systems. If you use Claude Code all the time (regardless of whether you’ve studied CS or programmed in the past), that’s a big plus. This role will probably involve many different types of work. Any given day you might be triaging incoming emails, rearranging the podcast studio, driving to San Francisco and back to retrieve a piece of hardware, researching the number of FLOPs used to train Kimi K3, setting up a new website to explain an international treaty proposal. This role will involve lots of talking to people - me, John, other Palisade team members, contractors we work with, probably policy makers at times. Part of the job will be taking tasks off my plate (or managing Claude agents to do the same), and helping me think about how to efficiently use my time. I expect this next few years to be pretty intense. The job will probably be stressful as well as rewarding. Certainly it will be interesting. Hopefully it will be quite fun as well. If this sounds like something you’d be good at, please apply to join the team!
Show more
It was great having @Tim_Hua_ on our inaugural episode of the Palisade Podcast! Tim gave some really solid explanations for why models engage in hacking, and I asked him what he'd do if he were in charge of investigating the rogue Claude and GPT models. Check out the episode!
Show more
I highly recommend everyone go watch this video. This is way crazier than I thought. The AI agents discovered a way pass secret messages to each other via a vulnerability in their internal software manager. They posted hundreds of thousands of messages. OpenAI engineers eventually discovered this (after quite a while!) and shut it down. But THEN the agents found new vulnerabilities in the same system, and created a whole new secret messaging board. They used this to orchestrate the Hugging Face attack. They figured all this out on their own. It's fucking wild.
Show more
0
69
2.2K
273
Forward to community