World models have emerged as one of the biggest directions in physical AI. At the same time, RL fine-tuning is unlocking capabilities in frontier models beyond what pretraining can achieve on its own
Can we get the best of both worlds?
We propose Q-Learning with World Models (QWM)
(1/7)