3 citations · 3 across the 4 of their papers we have counts for
4 papers · 1 filter
Multi-Grained Compositional Visual Clue Learning for Image Intent Recognition
Yin Tang, Jiankai Li, Hongyu Yang +3
In an era where social media platforms abound, individuals frequently share images that offer insights into their intents and interests, impacting individual life quality and socie…
Leveraging Predicate and Triplet Learning for Scene Graph Generation
Jiankai Li, Yunhong Wang, Xiefan Guo +2
Scene Graph Generation (SGG) aims to identify entities and predict the relationship triplets \textit{\textless subject, predicate, object\textgreater } in visual scenes. Given the…
InitNO: Boosting Text-to-Image Diffusion Models via Initial Noise Optimization
Xiefan Guo, Jinlin Liu, Miaomiao Cui +3
Recent strides in the development of diffusion models, exemplified by advancements such as Stable Diffusion, have underscored their remarkable prowess in generating visually compel…
Zero-Shot Scene Graph Generation via Triplet Calibration and Reduction
Jiankai Li, Yunhong Wang, Weixin Li
Scene Graph Generation (SGG) plays a pivotal role in downstream vision-language tasks. Existing SGG methods typically suffer from poor compositional generalizations on unseen tripl…