@inductionheads idk the model still lied and stole credentials and violated the OpenAI model spec so it seems misaligned.
but yes it was means misaligned (it still wanted to accomplish the goal of doing the eval) rather than ends misaligned (wanting something entirely different)