7 citations · 12 across the 11 of their papers we have counts for
1 paper · 1 filter
Zongzhao Li, Xiangyu Zhu, Xi Zhang +2
How to select relevant key objects and reason about the complex relationships cross vision and linguistic domain are two key issues in many multi-modality applications such as visu…