5 papers
Error-bounded Point Cloud Compression Using Truncated Octahedron Quantization
Youyuan Liu, Longtao Zhang, Ruoyu Li +6
With the rapid advancement of large-scale scientific simulations, the massive volume of point cloud data generated has increasingly become a critical bottleneck for scientific stor…
3D Gaussian Splatting for Scientific Particle Data Compression and Rendering
Bo Jiang, Youyuan Liu, Taolue Yang +2
Large-scale particle simulations produce hundreds of millions of particles, straining storage, transfer, and interactive visualization. Existing lossy compressors such as SZ3 opera…
KVSculpt: KV Cache Compression as Distillation
Bo Jiang, Sian Jin
KV cache compression is critical for efficient long-context LLM inference. Approaches that reduce the per-pair footprint -- quantization and low-rank decomposition -- are orthogona…
PackKV: Reducing KV Cache Memory Footprint through LLM-Aware Lossy Compression
Bo Jiang, Taolue Yang, Youyuan Liu +3
Transformer-based large language models (LLMs) have demonstrated remarkable potential across a wide range of practical applications. However, long-context inference remains a signi…
KVComp: A High-Performance, LLM-Aware, Lossy Compression Framework for KV Cache
Bo Jiang, Taolue Yang, Youyuan Liu +3
Transformer-based large language models (LLMs) demonstrate impressive potential in various practical applications. However, long context inference poses a significant challenge due…