145 citations · 145 across the 7 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
UniVL: Unified Vision-Language Embedding for Spatially Grounded Contextual Image Generation
Jiayun Wang, Yu Wang, Weijie Gan +2
We introduce spatially grounded contextual image generation, a controllable image generation task that reframes the conditioning paradigm. Instead of supplying a reference image an…
cs.CV2026
Dual-Prompt CLIP with Hybrid Visual Encoders for Occluded Person Re-Identification
Zhangjian Ji, Shaotong Qiao, Kai Feng +1
Occluded person re-identification focuses on matching partially visible pedestrians across multiple camera views. However, occlusions disrupt body-region cues, thereby complicating…