6 citations · 7 across the 5 of their papers we have counts for
5 papers
"See the World, Discover Knowledge": A Chinese Factuality Evaluation for Large Vision Language Models
Jihao Gu, Yingyao Wang, Pi Bu +17
The evaluation of factual accuracy in large vision language models (LVLMs) has lagged behind their rapid development, making it challenging to fully reflect these models' knowledge…
Video Referring Expression Comprehension via Transformer with Content-conditioned Query
Ji Jiang, Meng Cao, Tengtao Song +3
Video Referring Expression Comprehension (REC) aims to localize a target object in videos based on the queried natural language. Recent improvements in video REC have been made usi…
Improve Retrieval-based Dialogue System via Syntax-Informed Attention
Tengtao Song, Nuo Chen, Ji Jiang +2
Multi-turn response selection is a challenging task due to its high demands on efficient extraction of the matching features from abundant information provided by context utterance…
A Dynamic Graph Interactive Framework with Label-Semantic Injection for Spoken Language Understanding
Zhihong Zhu, Weiyuan Xu, Xuxin Cheng +2
Multi-intent detection and slot filling joint models are gaining increasing traction since they are closer to complicated real-world scenarios. However, existing approaches (1) foc…
Video Referring Expression Comprehension via Transformer with Content-aware Query
Ji Jiang, Meng Cao, Tengtao Song +1
Video Referring Expression Comprehension (REC) aims to localize a target object in video frames referred by the natural language expression. Recently, the Transformerbased methods…