We built a simple security layer for agentic infrastructure.
When an attack hits, we don’t let it touch production. We capture it, spin up a sandbox with the same infra pattern, and let the agent learn there.
Same architecture. Different risk level.
The sandbox becomes a training ground: replay the attack → map the failure path → test the fix → store the lesson.
Security for agents can’t just be a wall. Agents write code, call tools, and touch APIs. The next layer has to be a simulator, a teacher, and a memory system.
Attack → sandbox → learn → stronger infra.
Would love thoughts from people building agent runtimes, evals, and AI infra:
@hackgoofer @neutralino1 @JinjingLiang @xdotli @mo_tiwari @typesafeai @wandb @agihouse_org
@ktaletsk @AhmadMustafaAn1 @infiniter3grets @mehular0ra @RebbaVenkatarao
@typesafeai,
@benchflow_aiMo @AGIHouseSFAhmad @mehular0ratwo
@shvarts_eugene,
@EShvartsman.
@langchain
@LangChainAI
@e2b_dev
@modal
@e2bdev
@swyx