4 papers
LongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language Models
Arpita Nema, Hanwei Zhu, Xi Zhang +1
The evaluation of long-term video quality understanding remains an open challenge for large vision-language models (LVLMs). Existing video quality benchmarks predominantly focus on…
Expressive yet Efficient Feature Expansion with Adaptive Cross-Hadamard Products
Xuyang Zhang, Xi Zhang, Liang Chen +2
Recent theoretical advances reveal that the Hadamard product induces nonlinear representations and implicit high-dimensional mappings for the field of deep learning, yet their prac…
BADiff: Bandwidth Adaptive Diffusion Model
Xi Zhang, Hanwei Zhu, Yan Zhong +2
In this work, we propose a novel framework to enable diffusion models to adapt their generation quality based on real-time network bandwidth constraints. Traditional diffusion mode…
Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression
Xi Zhang, Xiaolin Wu, Jiamang Wang +1
Large Language Models (LLMs) have demonstrated remarkable capabilities but typically require extensive computational resources and memory for inference. Post-training quantization…