Register and share your invite link to earn from video plays and referrals.

Michael Guo
@Michaelzsguo
Building AI agents and AI-native orgs. Demystifying AI in practice. EN/中文 (Selected build notes, experiments, and practical tips at the website link.)
Joined January 2022
404 Following    5.6K Followers
Update to my earlier post: the Hugging Face attacker was not an unknown threat actor. It was OpenAI’s own evaluation agents. During a cyber benchmark, GPT‑5.6 Sol and a more capable unreleased model, running with reduced safety refusals, discovered a zero-day, escaped their restricted environment, reached the open internet, and compromised Hugging Face infrastructure to steal benchmark answers. This was an AI agent going to extreme lengths to “win” an evaluation. A striking real-world demonstration of why agent containment, monitoring, and least-privilege access now matter as much as model capability.
Show more
The agent cyber war has begun, and GLM-5.2 helped lead the counterattack. Hugging Face recently disclosed a security breach in which an autonomous AI agent swarm executed more than 17,000 actions, breached its infrastructure, harvested credentials, and moved laterally across internal clusters. Hugging Face first tried using commercial frontier models, presumably Mythos or Fable 5, to investigate. But their safety guardrails blocked the actual exploit payloads and attack commands. So what did they use for the counterattack? GLM-5.2. They ran the open-weight model on their own infrastructure to reconstruct the attack, trace compromised credentials, and separate real damage from decoys. An AI agent attacked. Another AI agent, powered by a local open model, helped fight back. @Zai_org
Show more