CORTEXA
← Browse

Fangzhao Wu

2 papers indexed

arxivcs.HC2026-07-11

Learning behavior accounts for background-related advantage in AI-assisted education

Jingwei Yi, Yueqi Xie, Jiyan He, Rui Ye, Junming Huang, Bin Zhu, et al.

Generative AI has been found, and will likely be found increasingly, useful in education. However, existing AI-for-education studies provide inconsistent evidence on its average effects. More broadly, research on prior educational technologies shows that average effects often mas…

View free PDFSource page
arxivcs.AIcs.CR2026-07-01

HARC: Coupling Harmfulness and Refusal Directions for Robust Safety Alignment

Shei Pern Chua, Hao Wu, Qianli Ma, Fangzhao Wu

Understanding how aligned LLMs internally represent safety is critical for diagnosing alignment vulnerabilities, as it explains why jailbreaks succeed and informs the design of robust alignment strategies. Prior work shows that aligned LLMs encode harmfulness and refusal as separ…

View free PDFSource page