14 papers
Spend Bits Where Queries Look: KV Cache Vector Quantization with Attention-Preserving Transforms
Samuel Fernández-Menduiña, Amir Ziashahabi, Eduardo Pavez +2
Long-context LLM decoding reads the key-value (KV) cache at every step. Loading it takes longer than computing attention over it, so throughput is bandwidth-bound. Hence, reducing…
Motion Estimation Techniques for Volumetric Video Attribute Compression
Haoran Hong, Eduardo Pavez, Antonio Ortega +2
Point cloud compression relies on techniques to compress both geometry and attributes. Motion-based approaches for dynamic solid point cloud geometry compression within the geometr…
GS-NFS: Bandwidth-adaptive Streaming of Dynamic Gaussian Splats and Point Clouds
Rajrup Ghosh, Haodong Wang, Haoran Hong +6
Dynamic 3D Gaussian Splatting (3DGS) holds great promise as a 3D video streaming technology since it can represent complex 3D scenes with high fidelity. In this approach, every fra…
L2G-Net: Local to Global Spectral Graph Neural Networks via Cauchy Factorizations
Samuel Fernández-Menduiña, Eduardo Pavez, Antonio Ortega
Despite their theoretical advantages, spectral methods based on the graph Fourier transform (GFT) are seldom used in graph neural networks (GNNs) due to the cost of computing the e…
FaSST: Fast Sparsifying Secondary Transform
Darukeesan Pakiyarajah, Samuel Fernández-Menduiña, Eduardo Pavez +2
Data-dependent secondary transforms, which aim to decorrelate coefficients of a separable primary transform, can improve residual coding efficiency; however, their deployment is of…
Rate-Distortion Optimization for Ensembles of Non-Reference Metrics
Xin Xiong, Samuel Fernández-Menduiña, Eduardo Pavez +3
Non-reference metrics (NRMs) can assess the visual quality of images and videos without a reference, making them well-suited for the evaluation of user-generated content. Nonethele…