The real world is multimodal. For AI to understand and recreate it, models need to learn across modalities.
In our latest blog, we show how Miles supports that learning with a shared post-training design for VLMs and diffusion models.
Link in the comments. đ