Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Crowdsource, Crawl, or Generate? Creating SEA-VL, a Multicultural Vision-Language Dataset for Southeast Asia
Samuel Cahyawijaya, Holy Lovenia, Joel Ruben Antony Moniz +89
Southeast Asia (SEA) is a region of extraordinary linguistic and cultural diversity, yet it remains significantly underrepresented in vision-language (VL) research. This often resu…
cs.CV2024
Efficient Human-Object-Interaction (EHOI) Detection via Interaction Label Coding and Conditional Decision
Tsung-Shan Yang, Yun-Cheng Wang, Chengwei Wei +2
Human-Object Interaction (HOI) detection is a fundamental task in image understanding. While deep-learning-based HOI methods provide high performance in terms of mean Average Preci…