2 citations · 4 across the 4 of their papers we have counts for
4 papers
Knowledge Condensation and Reasoning for Knowledge-based VQA
Dongze Hao, Jian Jia, Longteng Guo +8
Knowledge-based visual question answering (KB-VQA) is a challenging task, which requires the model to leverage external knowledge for comprehending and answering questions grounded…
Towards Efficient and Effective Text-to-Video Retrieval with Coarse-to-Fine Visual Representation Learning
Kaibin Tian, Yanhua Cheng, Yi Liu +3
In recent years, text-to-video retrieval methods based on CLIP have experienced rapid development. The primary direction of evolution is to exploit the much wider gamut of visual a…
Cross-view Semantic Alignment for Livestreaming Product Recognition
Wenjie Yang, Yiyi Chen, Yan Li +4
Live commerce is the act of selling products online through live streaming. The customer's diverse demands for online products introduce more challenges to Livestreaming Product Re…
Cross-Domain Product Representation Learning for Rich-Content E-Commerce
Xuehan Bai, Yan Li, Yanhua Cheng +3
The proliferation of short video and live-streaming platforms has revolutionized how consumers engage in online shopping. Instead of browsing product pages, consumers are now turni…