3 papers
cs.CL2026
NestedKV: Nested Memory Routing for Long-Context KV Cache Compression
Hong Chen, Xiang Liu, Yubo Gao +5
Long-context language models are limited by the memory footprint of the key-value (KV) cache. Existing training-free KV compression methods usually rank tokens by one importance si…
cs.LG2026
DP-SelFT: Differentially Private Selective Fine-Tuning for Large Language Models
Haichao Sha, Zihao Wang, Yuncheng Wu +2
Large language models (LLMs) are commonly adapted to downstream tasks through fine-tuning, but fine-tuning data often contains sensitive information that may be leaked by the resul…
cs.LG2026
Federated Nested Learning: Collaborative Training of Self-Referential Memories for Test-Time Adaptation
Hong Chen, Pengcheng Wu, Yuanguo Lin +4
We rethink Federated Learning (FL) from a nested learning perspective, framing the core challenge as how to collaboratively learn optimization rules, not just static models, to tac…