Kind of wondersome if they will start to combine LLM generation with actual world models/GPU engines. Imagine that your model generates the wireframes, the Engine generates the physics and then another model fills in the actual visuals, and gaps... So you have realistic physics and gaps are filled in... Will also help with image retention more, if objects moved behind each other.