2 papers
cs.CL2026
ToxiTrace: Gradient-Aligned Training for Explainable Chinese Toxicity Detection
Boyang Li, Hongzhe Shou, Yuanyuan Liang +2
Existing Chinese toxic content detection methods mainly target sentence-level classification but often fail to provide readable and contiguous toxic evidence spans. We propose \tex…
cs.LG2025
Sparse Attention across Multiple-context KV Cache
Ziyi Cao, Qingyi Si, Jingbin Zhang +1
Large language models face significant cost challenges in long-sequence inference. To address this, reusing historical Key-Value (KV) Cache for improved inference efficiency has be…