Our vision is for AI that uses world models to adapt in new and dynamic environments and efficiently learn new skills.
We’re sharing V-JEPA 2, a new world model with state-of-the-art performance in visual understanding and prediction.
V-JEPA 2 is a 1.2 billion-parameter model, trained on video, that can enable zero-shot planning in robots—allowing them to plan and execute tasks in unfamiliar environments.
Learn more about V-JEPA 2 ➡️
As we continue working toward our goal of achieving advanced machine intelligence (AMI), we’re also releasing three new benchmarks for evaluating how well existing models can reason about the physical world from video.
Learn more and download the new benchmarks ➡️