15 papers
Spend Bits Where Queries Look: KV Cache Vector Quantization with Attention-Preserving Transforms
Samuel Fernández-Menduiña, Amir Ziashahabi, Eduardo Pavez +2
Long-context LLM decoding reads the key-value (KV) cache at every step. Loading it takes longer than computing attention over it, so throughput is bandwidth-bound. Hence, reducing…
Reduced-complexity Adaptive Loop Filtering via Input-dependent Graph Filters
Wen-Yang Lu, Eduardo Pavez, Antonio Ortega +2
Adaptive Loop Filtering is an important tool for suppressing compression artifacts in modern video codecs. In the enhanced compression model (ECM), a software test model used for e…
Motion Estimation Techniques for Volumetric Video Attribute Compression
Haoran Hong, Eduardo Pavez, Antonio Ortega +2
Point cloud compression relies on techniques to compress both geometry and attributes. Motion-based approaches for dynamic solid point cloud geometry compression within the geometr…
GS-NFS: Bandwidth-adaptive Streaming of Dynamic Gaussian Splats and Point Clouds
Rajrup Ghosh, Haodong Wang, Haoran Hong +6
Dynamic 3D Gaussian Splatting (3DGS) holds great promise as a 3D video streaming technology since it can represent complex 3D scenes with high fidelity. In this approach, every fra…
FaSST: Fast Sparsifying Secondary Transform
Darukeesan Pakiyarajah, Samuel Fernández-Menduiña, Eduardo Pavez +2
Data-dependent secondary transforms, which aim to decorrelate coefficients of a separable primary transform, can improve residual coding efficiency; however, their deployment is of…
Rate-Distortion Optimization for Ensembles of Non-Reference Metrics
Xin Xiong, Samuel Fernández-Menduiña, Eduardo Pavez +3
Non-reference metrics (NRMs) can assess the visual quality of images and videos without a reference, making them well-suited for the evaluation of user-generated content. Nonethele…