CORTEXA
← Browse

Tiantong Wang

1 paper indexed

arxivcs.LGcs.AI2026-07-20

CoCurve: Cross-Module Co-Pruning Curvature for Training-Free Structured LLM Pruning

Zhiren Gong, Zihao Zeng, Zijie Wang, Tiantong Wang, Chau Yuen, Wei Yang Bryan Lim

Structured pruning compresses large language models (LLMs) by removing whole computational units, such as attention heads and feed-forward (FFN) channel groups. Most training-free methods, however, rank these units independently, implicitly treating the loss from pruning a set as…

View free PDFSource page