Federated fine-tuning of large pre-trained models increasingly relies on Low-Rank Adaptation (LoRA) to reduce communication and computation, but heterogeneous clients can make adapter aggregation unstable. We identify the data-parameter interference as a geometric source of this…
The key-value (KV) cache has become a first-order memory object in LLM serving rather than a temporary per-request tensor. This survey classifies more than thirty KV-management systems and frameworks using four axes: locality, lifetime, ownership, and substrate. The axes reveal f…
The research examines how digital inclusive finance reshapes the rural labor market using an auditable index system and an interpretable learning pipeline. We construct a four-pillar framework for the rural labor market covering labor behavior, labor structure, security and fairn…