Showing 2026Show all
2 papers · 1 filter
cs.AI2026
FAST-GOAL: Fast and Efficient Global-local Object Alignment Learning
Hyungyu Choi, Young Kyun Jang, Chanho Eom
Vision-language models such as CLIP have shown impressive capabilities in aligning images and text, but they often struggle with lengthy and detailed text descriptions due to pre-t…
cs.CV2026
DiCo: Disentangled Concept Representation for Text-to-image Person Re-identification
Giyeol Kim, Chanho Eom
Text-to-image person re-identification (TIReID) aims to retrieve person images from a large gallery given free-form textual descriptions. TIReID is challenging due to the substanti…