61 citations · 112 across the 18 of their papers we have counts for
18 papers
MIT-10M: A Large Scale Parallel Corpus of Multilingual Image Translation
Bo Li, Shaolin Zhu, Lijie Wen
Image Translation (IT) holds immense potential across diverse domains, enabling the translation of textual content within images into various languages. However, existing datasets…
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
Junzhe Chen, Tianshu Zhang, Shiyu Huang +4
Despite the recent breakthroughs achieved by Large Vision Language Models (LVLMs) in understanding and responding to complex visual-textual contexts, their inherent hallucination t…
CINet: Realizing Incremental Trajectory Prediction with Prior-Aware Continual Causal Intervention
Xiaohe Li, Feilong Huang, Zide Fan +4
Trajectory prediction for multi-agents in complex scenarios is crucial for applications like autonomous driving. However, existing methods often overlook environmental biases, whic…
On the Robustness of Document-Level Relation Extraction Models to Entity Name Variations
Shiao Meng, Xuming Hu, Aiwei Liu +4
Driven by the demand for cross-sentence and large-scale relation extraction, document-level relation extraction (DocRE) has attracted increasing research interest. Despite the cont…
FSMR: A Feature Swapping Multi-modal Reasoning Approach with Joint Textual and Visual Clues
Shuang Li, Jiahua Wang, Lijie Wen
Multi-modal reasoning plays a vital role in bridging the gap between textual and visual information, enabling a deeper understanding of the context. This paper presents the Feature…
ChatCite: LLM Agent with Human Workflow Guidance for Comparative Literature Summary
Yutong Li, Lu Chen, Aiwei Liu +2
The literature review is an indispensable step in the research process. It provides the benefit of comprehending the research problem and understanding the current research situati…