3 papers
cs.CL2026
TopoCompress: Long Context Compression via Graph-Wired Semantic Trajectories
Daniel Agyei Asante, Yang Li
Long-context compression is essential for reducing the cost and latency of large language model inference. However, existing methods can fragment important evidence, require additi…
cs.LG2025
Decomposed Trust: Privacy, Adversarial Robustness, Ethics, and Fairness in Low-Rank LLMs
Daniel Agyei Asante, Md Mokarram Chowdhury, Yang Li
Large language models (LLMs) have driven major advances across domains, yet their massive size hinders deployment in resource-constrained settings. Low-rank factorization addresses…
cs.LG2025
IMPACT: Importance-Aware Activation Space Reconstruction
Md Mokarram Chowdhury, Daniel Agyei Asante, Ernie Chang +1
Large language models (LLMs) achieve strong performance across diverse domains but remain difficult to deploy in resource-constrained environments due to their size. Low-rank compr…