What's interesting is that Astra and Fable are RLed so much for agentic coding that their English language skills have really regressed. It somewhat makes sense.
RLing heavily on coding data makes you worse at communication in plain English, as you basically learned to speak code very well.
Most likely they will have to modify their RL envs to penalize this poor prose and/or add new RL envs to teach the model to have better prose.