2 papers
cs.AI2026
ChartAnchor: Chart Grounding with Structural-Semantic Fidelity
Xinhang Li, Jingbo Zhou, Pengfei Luo +2
Recent advances in multimodal large language models (MLLMs) highlight the need for benchmarks that rigorously evaluate structured chart comprehension. Chart grounding refers to the…
cs.IR2025
ImageScope: Unifying Language-Guided Image Retrieval via Large Multimodal Model Collective Reasoning
Pengfei Luo, Jingbo Zhou, Tong Xu +3
With the proliferation of images in online content, language-guided image retrieval (LGIR) has emerged as a research hotspot over the past decade, encompassing a variety of subtask…