WorldDirector proposes a controllable video world-model framework that separates semantic motion orchestration from visual rendering. An LLM coordinates 3D trajectories for objects and camera movements, and these trajectories are then used as control signals for video generation. According to the abstract, this design improves physical consistency and preserves the visual identity of dynamic entities when they re-enter the scene after being absent for extended periods. The framework is intended to support complex, long-duration events and unrestricted viewpoint exploration. The supplied project page provides additional implementation context, but the abstract alone does not establish the full evaluation scope or comparative results.
No heat snapshots are available in the last 24 hours.