3 papers
cs.CV2025
GloTok: Global Perspective Tokenizer for Image Reconstruction and Generation
Xuan Zhao, Zhongyu Zhang, Yuge Huang +6
Existing state-of-the-art image tokenization methods leverage diverse semantic features from pre-trained vision models for additional supervision, to expand the distribution of lat…
cs.CV2025
Switchable Token-Specific Codebook Quantization For Face Image Compression
Yongbo Wang, Haonan Wang, Guodong Mu +7
With the ever-increasing volume of visual data, the efficient and lossless transmission, along with its subsequent interpretation and understanding, has become a critical bottlenec…
cs.CV2025
From Enhancement to Understanding: Build a Generalized Bridge for Low-light Vision via Semantically Consistent Unsupervised Fine-tuning
Sen Wang, Shao Zeng, Tianjun Gu +8
Low-level enhancement and high-level visual understanding in low-light vision have traditionally been treated separately. Low-light enhancement improves image quality for downstrea…