CORTEXA
← Browse

Chenming Zhu

1 paper indexed

arxivcs.CV2026-07-04

G$^2$TAM: Geometry Grounded Track Anything Model

Chenming Zhu, Peizhou Cao, Jingli Lin, Wenbo Hu, Yunlong Ran, Jiangmiao Pang, et al.

Human spatial understanding arises from jointly perceiving geometry and semantics, enabling consistent object identification and localization across viewpoints and time. Current video segmentation models depend on explicit object appearance memory banks for instance tracking, yet…

View free PDFSource page