Most modern video diffusion models treat time as a symmetric axis, attending to past and future alike, yet we keep calling them world models and asking them to predict forward. This post is my late-night attempt to understand what it would mean for such a model to have an arrow of time, why variational free energy already has one baked in, and why a recent paper had to perform philosophical surgery on Sora-class models to make them interactive.