CORTEXA
← Browse

Gaowen Liu

3 papers indexed

arxivcs.CLcs.LG2026-07-21

Stochastic Meta-Unlearning: Bridging Language Backbone and Multimodal Unlearning

Zijie Liu, Jinhao Duan, Gaowen Liu, Sijia Liu, Tianlong Chen

Machine unlearning for vision-language models (VLMs) remains underexplored. Unlike language models, VLMs combine a language backbone with visual components, which makes unlearning more complex. There is a surprising phenomenon when moving from single-modality unlearning to VLM un…

View free PDFSource page
arxivcs.LG2026-07-15

The Hyperspherical Geometry of CLIP Latent Space: A Semantic Mixture Model

Zijie Yu, Gaowen Liu, Ramana Rao Kompella, Philip S. Yu, Yue Song

Contrastive Language-Image Pretraining (CLIP) representations form a semantic embedding space governed by cosine similarity, reflecting an intrinsic hyperspherical geometry. However, existing probabilistic interpretations typically rely on Gaussian assumptions, which fail to capt…

View free PDFSource page
arxivcs.ROcs.AI2026-06-26

Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models?

Guoheng Sun, Kaixi Feng, Shwai He, Xiaochuan Gong, Yexiao He, Ziyao Wang, et al.

Vision-Language-Action (VLA) models enable instruction-driven robotic manipulation, but they inherit oversized language backbones from pretrained VLMs whose capacity far exceeds what is needed for short robotic instructions. This raises a basic question: how much of a VLA model i…

View free PDFSource page