Skip to content

Author

Zi-Li Yi

We have 4 of 25 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

SynVAR: Synergizing Spatial and Semantic Alignment in Visual Autoregressive Model

VAR has gained widespread popularity due to its next-scale prediction paradigm. However, it faces substantial performance bottlenecks when handling complex scenes with multiple objects and attributes. Existing diffusion-based enhancement methods fail to adequately address the unique challenge of cross-scale error propa...

Zhen-Nan Chen, Tianxing Shi, Pengcheng Xu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

VideoX-Qwen: Data-Centric Instruction-Based Video Editing

Progress in general-purpose video editing depends on constructing large-scale paired supervision and effectively adapting video-generation backbones to instruction-driven editing. Unlike video generation, video editing must execute a requested transformation while preserving unrelated subjects, scene structure, motion,...

J.Jenny Li, Di Shao, Xin-Yu Chen et al. · 0 citations
Preprint Aug 2026

InstructVVT: Instruction-Driven Video Virtual Try-On without Auxiliary Spatial Priors

InstructVVT is proposed, an instruction-driven and reference-guided video virtual try-on framework based on a Diffusion Transformer that operates without inference-time spatial priors that outperforms state-of-the-art open-source methods in garment fidelity, structural preservation, and temporal consistency, despite re...

Di Shao, Song-Han Wu, Xin-Yu Chen et al. · 0 citations
Preprint Aug 2026

EgoMonth: A Month-Level Egocentric Video Benchmark for Long-Term Spatiotemporal Memory

Evaluation of state-of-the-art open-source and closed-source MLLMs reveals that current MLLMs function as lossy summarizers rather than faithful memorizers, highlighting the need for architectures with genuine long-term spatiotemporal memory.

Weitao Chen, Jiaxing Hu, Xie Tianyidan et al. · 2 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.