CORTEXA
← Browse

Wei Gao

4 papers indexed

arxivcs.CV2026-07-21

Wave2Body: Rethinking mmWave Human Pose Estimation as Radar-to-Body Token Translation

Bo Liang, Chen Gong, Wei Gao, Chenren Xu

Millimeter-wave (mmWave) radar enables privacy-friendly human sensing, but its sparse point clouds are physical measurements of view-dependent electromagnetic reflections and only indirectly characterize body articulation. Recovering a complete 3D pose from such partial, geometry…

View free PDFSource page
arxivcs.CV2026-07-20

VGOcc: Learning Visual-Geometric Gaussians for Vision-Centric 3D Driving Occupancy Prediction

Junhong Lin, Xianda Guo, Kangli Wang, Yuqi Ye, Xiaoyu Liang, Yanlun Peng, et al.

Vision-only occupancy prediction requires recovering a semantic 3D occupancy field from calibrated surround-view images, where each view provides observations with ambiguous depth along camera rays. Existing methods have progressed from dense structured representations to sparse…

View free PDFSource page
arxivcs.CV2026-07-11

CoSAG: Compact Semantic Anchor Gaussians via Training-Free Rate-Distortion Coding

Yuang Jia, Jinlong Wang, Junhong Lin, Ruiting Dai, Wei Gao

Open-vocabulary 3D scene understanding is commonly achieved by embedding 2D vision-language features such as CLIP into a 3D Gaussian Splatting scene, turning it into a text-queryable semantic field. However, attaching a high-dimensional feature to each of millions of Gaussians in…

View free PDFSource page
arxivcs.CV2026-07-01

EFlow: Learning Evidence Flow for Long-Video Reasoning with Adaptive Reflection

Wenhao Zhang, Kuanwei Lin, Xuyi Yang, Wei Gao, Ge Li

Long-video reasoning is fundamentally constrained by how models acquire and utilize visual evidence. Existing tool-augmented video frameworks often interleave temporal grounding and answer reasoning within a single trajectory, causing early semantic hypotheses to bias evidence lo…

View free PDFSource page