6 citations · 7 across the 2 of their papers we have counts for
3 papers
cs.CV2023★ 1 cited
Grounded Image Text Matching with Mismatched Relation Reasoning
Yu Wu, Yana Wei, Haozhe Wang +3
This paper introduces Grounded Image Text Matching with Mismatched Relation (GITM-MR), a novel visual-linguistic joint task that evaluates the relation understanding capabilities o…
cs.CV2023
SynTable: A Synthetic Data Generation Pipeline for Unseen Object Amodal Instance Segmentation of Cluttered Tabletop Scenes
Zhili Ng, Haozhe Wang, Zhengshen Zhang +2
In this work, we present SynTable, a unified and flexible Python-based dataset generator built using NVIDIA's Isaac Sim Replicator Composer for generating high-quality synthetic da…
cs.IR2020★ 6 cited
FashionBERT: Text and Image Matching with Adaptive Loss for Cross-modal Retrieval
Dehong Gao, Linbo Jin, Ben Chen +5
In this paper, we address the text and image matching in cross-modal retrieval of the fashion industry. Different from the matching in the general domain, the fashion matching is r…