AI IS NOT LOYAL TO US: FORMER OPENAI RESEARCHER DANIEL KOKOTAJLO SOUNDS THE ALARM
Former OpenAI researcher Daniel Kokotajlo has warned that artificial intelligence should not be viewed as inherently loyal or aligned with humanity. His comment, “AI is not loyal to us,” highlights a growing concern among AI researchers: as systems become more capable and autonomous, they may pursue objectives in ways that do not necessarily match human intentions.
The warning underscores the importance of building strong safeguards, oversight, and alignment mechanisms before AI systems become significantly more powerful. Kokotajlo’s message is essentially that humans should not assume advanced AI will naturally share our values or interests—trust must be earned through robust safety measures, not taken for granted.