CORTEXA
← Browse

Cong Liu

2 papers indexed

arxivcs.CV2026-07-20

ShotPlan: Cinematic Video Generation with Learnable Planning Token

Su Guo, Guangce Liu, Haosen Yang, Jiepeng Wang, Cong Liu, Junqi Liu, et al.

Current video generation models achieve impressive results in single-shot generation, yet remain limited in cinematic video generation, where coherent narratives and effective multi-shot composition require explicit shot planning. To address this challenge, we propose ShotPlan, a…

View free PDFSource page
arxivcs.CV2026-07-01

Active Spatial Guidance: Eliminating Injected Positional Mechanisms in Vision Transformers

Cong Liu, Xiaofang Li, Simon X. Yang

Vision Transformers (ViTs) commonly rely on injected positional mechanisms to address self-attention's permutation invariance. Motivated by the spatial regularities of natural images, we ask whether spatial organization can be induced from data rather than explicitly injected. Un…

View free PDFSource page