Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Improving vision-language alignment with graph spiking hybrid Networks
Siyu Zhang, Wenzhe Liu, Yeming Chen +3
To bridge the semantic gap between vision and language (VL), it is necessary to develop a good alignment strategy, which includes handling semantic diversity, abstract representati…
cs.CV2024
Superpixel Semantics Representation and Pre-training for Vision-Language Task
Siyu Zhang, Yeming Chen, Yaoru Sun +4
The key to integrating visual language tasks is to establish a good alignment strategy. Recently, visual semantic representation has achieved fine-grained visual understanding by d…