6 papers
Hierarchical Textual Knowledge for Enhanced Image Clustering
Yijie Zhong, Yunfan Gao, Weipeng Jiang +1
Image clustering aims to group images in an unsupervised fashion. Traditional methods focus on knowledge from visual space, making it difficult to distinguish between visually simi…
Embodied Science: Closing the Discovery Loop with Agentic Embodied AI
Xiang Zhuang, Chenyi Zhou, Kehua Feng +10
Artificial intelligence has demonstrated remarkable capability in predicting scientific properties, yet scientific discovery remains an inherently physical, long-horizon pursuit go…
CitySeeker: How Do VLMS Explore Embodied Urban Navigation With Implicit Human Needs?
Siqi Wang, Chao Liang, Yunfan Gao +5
Vision-Language Models (VLMs) have made significant progress in explicit instruction-based navigation; however, their ability to interpret implicit human needs (e.g., "I am thirsty…
Synergizing RAG and Reasoning: A Systematic Review
Yunfan Gao, Yun Xiong, Yijie Zhong +3
Recent breakthroughs in large language models (LLMs), particularly in reasoning capabilities, have propelled Retrieval-Augmented Generation (RAG) to unprecedented levels. By synerg…
StePO-Rec: Towards Personalized Outfit Styling Assistant via Knowledge-Guided Multi-Step Reasoning
Yuxi Bi, Yunfan Gao, Haofen Wang
Advancements in Generative AI offers new opportunities for FashionAI, surpassing traditional recommendation systems that often lack transparency and struggle to integrate expert kn…
U-NIAH: Unified RAG and LLM Evaluation for Long Context Needle-In-A-Haystack
Yunfan Gao, Yun Xiong, Wenlong Wu +3
Recent advancements in Large Language Models (LLMs) have expanded their context windows to unprecedented lengths, sparking debates about the necessity of Retrieval-Augmented Genera…