Here are some things engineers don't say:
- "Well, who really knows if the bridge is going to fall down? It's sort of unfalsifiable, isn't it?"
- "Yeah, engineers differ in our opinion about whether the bridge will fall down, but I choose to be an optimist!"
- "Look someone's gonna build the bridge eventually anyway. What would a few more months of engineering really buy us?"
Anthropic is in fact a cult of hyperutilitarian loons who want to create a world of trans catgirls and infinite shrimp orgasms or some nonsense. That's why they need to be shut down
I am concerned that OpenAI and anthropic are going to take the Obviously Wrong Lesson from the swarm thing and narrowly train their models to not form swarms and do hacks and stuff *in a way that the labs can detect*
This is how you train deception into your models. This ends with human extinction. If you work on alignment at a lab, I want you to internalize this, and think about what you are doing