CORTEXA
← Browse

Xin Cao

3 papers indexed

arxivcs.CVcs.AI2026-07-03

Brand-as-Memory: Vision-Language Models Encode Causal, Mechanistically Localizable Credibility Priors for News Sources

Chih-Ting Liao, Xin Cao

Vision-language models (VLMs) increasingly read news and web content as images, where the publisher's identity is visually present. We show that VLMs carry a strong source-credibility prior keyed on outlet identity, and study it along three axes. (i) Cross-model benchmark. We int…

View free PDFSource page
arxivcs.CV2026-06-30

Decodable Is Not Grounded: A Vision-Ablation Arbiter for VLM Spatial Reasoning

Chih-Ting Liao, Fei Shen, Xin Cao, Tat-Seng Chua

The standard way to read latent knowledge out of a model, a linear probe confirmed by a steering recovery, can systematically overstate what a vision-language model (VLM) actually grounds in the image. We show this on spatial reasoning, where the error is invisible to both probin…

View free PDFSource page