3 papers
cs.CL2026
Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation
Weiqing Luo, Zongye Hu, Xiao Wang +3
Visual evidence selection is a critical component of multimodal retrieval-augmented generation (RAG), yet existing methods typically rely on semantic relevance or surface-level sim…
cs.CV2025
Task-Aware Resolution Optimization for Visual Large Language Models
Weiqing Luo, Zhen Tan, Yifan Li +4
Real-world vision-language applications demand varying levels of perceptual granularity. However, most existing visual large language models (VLLMs), such as LLaVA, pre-assume a fi…
cs.IR2025
Benchmarking Recommendation, Classification, and Tracing Based on Hugging Face Knowledge Graph
Qiaosheng Chen, Kaijia Huang, Xiao Zhou +3
The rapid growth of open source machine learning (ML) resources, such as models and datasets, has accelerated IR research. However, existing platforms like Hugging Face do not expl…