Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Lang2Act: Fine-Grained Visual Reasoning through Self-Emergent Linguistic Toolchains
Yuqi Xiong, Chunyi Peng, Zhipeng Xu +6
Visual Retrieval-Augmented Generation (VRAG) enhances Vision-Language Models (VLMs) by incorporating external visual documents to address a given query. Existing VRAG frameworks us…
cs.AI2026
DIAL-KG: Schema-Free Incremental Knowledge Graph Construction via Dynamic Schema Induction and Evolution-Intent Assessment
Weidong Bao, Yilin Wang, Ruyu Gao +3
Knowledge Graphs (KGs) are foundational to applications such as search, question answering, and recommendation. Conventional knowledge graph construction methods are predominantly…
cs.AI2025
Benchmarking Retrieval-Augmented Generation in Multi-Modal Contexts
Zhenghao Liu, Xingsheng Zhu, Tianshuo Zhou +5
With the rapid advancement of Multi-modal Large Language Models (MLLMs), their capability in understanding both images and text has greatly improved. However, their potential for l…