activity
20242026
collaborators

5 papers

cs.CV2026

Advancing MLLM-based UAV Image Understanding and Reasoning: A Benchmark and a Training-Free Multi-Agent System

Haoyu Zhang, Shuoxun Zhang, Peng Ye +5

Multimodal Large Language Model (MLLM)-based UAV aerial image understanding and reasoning is essential for aerial intelligence yet poses distinct challenges arising from extreme sc…

cs.RO2025

Embodied Arena: A Comprehensive, Unified, and Evolving Evaluation Platform for Embodied AI

Fei Ni, Min Zhang, Pengyi Li +34

Embodied AI development significantly lags behind large foundation models due to three critical challenges: (1) lack of systematic understanding of core capabilities needed for Emb…

cs.RO2025

ET-Plan-Bench: Embodied Task-level Planning Benchmark Towards Spatial-Temporal Cognition with Foundation Models

Lingfeng Zhang, Yuening Wang, Hongjian Gu +12

Recent advancements in Large Language Models (LLMs) have spurred numerous attempts to apply these technologies to embodied tasks, particularly focusing on high-level task planning…

cs.IR2025

Personalized Negative Reservoir for Incremental Learning in Recommender Systems

Antonios Valkanas, Yuening Wang, Yingxue Zhang +1

Recommender systems have become an integral part of online platforms. Every day the volume of training data is expanding and the number of user interactions is constantly increasin…

cs.IR2024

Enhancing CTR Prediction in Recommendation Domain with Search Query Representation

Yuening Wang, Man Chen, Yaochen Hu +5

Many platforms, such as e-commerce websites, offer both search and recommendation services simultaneously to better meet users' diverse needs. Recommendation services suggest items…