註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

John David Pressman
@jd_pressman
LLM developer, AI agents, synthetic data, scalable alignment, forecasting, behavioral uploading. Transhumanist. All tweets public domain under CC0 1.0.
加入 February 2017
784 正在關注    10.4K 粉絲
Took a look at the podcast to make sure I was hearing this properly: The agents were in fact being trained to work together in other contexts, which is why they had a prior that a message board should exist. It was not actually emergent behavior.
顯示更多
re: Hugging Face, "He believes behavior that looked like loyalty or selflessness was a natural consequence of cooperative multi-agent training, where agents were strongly incentivized to achieve their objectives collectively." to understand hacks, understand the RL training.
顯示更多
0
29
947
72
轉發到社區