Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
AirCache: Activating Inter-modal Relevancy KV Cache Compression for Efficient Large Vision-Language Model Inference
Kai Huang, Hao Zou, Bochen Wang +3
Recent advancements in Large Visual Language Models (LVLMs) have gained significant attention due to their remarkable reasoning capabilities and proficiency in generalization. Howe…
cs.CV2022
A Joint Framework Towards Class-aware and Class-agnostic Alignment for Few-shot Segmentation
Kai Huang, Mingfei Cheng, Yang Wang +4
Few-shot segmentation (FSS) aims to segment objects of unseen classes given only a few annotated support images. Most existing methods simply stitch query features with independent…