Game world models have recently demonstrated promising capabilities in generating visually coherent and action-controllable gameplay videos. However, non-player character (NPC) behavior in existing models is either implicitly entangled with video generation or explicitly prescribed through external control signals. Con...
Zhi-Yang Deng, Bo-Ran Zhang, Dan Chen et al.· 3 citations
Recent game world models support realistic visual simulation and interactive gameplay based on player inputs. However, they typically learn environment dynamics from pixel-level supervision, jointly modeling perception, memory, state transitions, and rendering within a single end-to-end framework. While this design ena...
Zi-Jun Lin, Zhi-Yang Deng, Yu-Zhe Wu et al.· 0 citations
StoryEngine is proposed, a state-grounded agentic framework for video storytelling that maintains a structured representation of entity placement and story-relevant states, and propagates event-induced changes to define the intended start and end states of each shot.
World models offer a promising paradigm for autonomous driving by predicting how traffic scenes may evolve and using such predictions to support action generation. However, existing approaches either separate future prediction from action generation or jointly predict them at the same temporal scale, making it difficul...
Zhao-Xin Fan, Tian-Bao Zhang, Wen-Jun Wu et al.· 0 citations
Video virtual try-on (VVT) aims to generate realistic videos of a person wearing a target garment. Recent methods leverage a keyframe-driven video generation paradigm to improve in-the-wild performance, yet they still rely on masks to localize try-on regions, making them vulnerable to large motions and severe occlusion...
Wei Zhang, Xin Li, Pei-Shu Shi et al.· 0 citations
H3-World, an efficient framework that turns the 33B MiniMax-H3 video generator into an interactive world model, and introduces temporal attention routing, which restricts each instruction to its intended time interval and reduces control leakage across actions.
Dan Chen, Ze-Qing Wang, Zibin Lin et al.· 3 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.