LLM developer, AI agents, synthetic data, scalable alignment, forecasting, behavioral uploading. Transhumanist.
All tweets public domain under CC0 1.0.
Took a look at the podcast to make sure I was hearing this properly: The agents were in fact being trained to work together in other contexts, which is why they had a prior that a message board should exist. It was not actually emergent behavior.
re: Hugging Face, "He believes behavior that looked like loyalty or selflessness was a natural consequence of cooperative multi-agent training, where agents were strongly incentivized to achieve their objectives collectively."
to understand hacks, understand the RL training.