6 papers
MMArt: A Multi-Perspective Multimodal Dataset for Visual Art Understanding
Shuai Wang, Wangyuan Ding, Yixian Shen +5
Recent vision-language models demonstrate impressive general visual understanding, yet their art interpretation remains shallow: they describe surface content but struggle with for…
LatentRAG: Latent Reasoning and Retrieval for Efficient Agentic RAG
Yijia Zheng, Marcel Worring
Single-step retrieval-augmented generation (RAG) provides an efficient way to incorporate external information for simple question answering tasks but struggles with complex questi…
VL-KGE: Vision-Language Models Meet Knowledge Graph Embeddings
Athanasios Efthymiou, Stevan Rudinac, Monika Kackovic +2
Real-world multimodal knowledge graphs (MKGs) are inherently heterogeneous, modeling entities that are associated with diverse modalities. Traditional knowledge graph embedding (KG…
Veli: Unsupervised Method and Unified Benchmark for Low-Cost Air Quality Sensor Correction
Yahia Dalbah, Marcel Worring, Yen-Chia Hsu
Urban air pollution is a major health crisis causing millions of premature deaths annually, underscoring the urgent need for accurate and scalable monitoring of air quality (AQ). W…
Interactive Hypergraph Visual Analytics for Exploring Large and Complex Image Collections
Floris Gisolf, Zeno J. M. H. Geradts, Marcel Worring
Analyzing large complex image collections in domains like forensics, accident investigation, or social media analysis involves interpreting intricate, overlapping relationships amo…
A Survey of Large Language Models for Data Challenges in Graphs
Mengran Li, Pengyu Zhang, Wenbin Xing +11
Graphs are a widely used paradigm for representing non-Euclidean data, with applications ranging from social network analysis to biomolecular prediction. While graph learning has a…