7 papers
CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection
Huidong Feng, Wentao Chen, Jie Chen +5
With the rapid advancement of artificial intelligence generated content (AIGC) technologies, video forgery has become increasingly prevalent, posing new challenges to public discou…
Reasoning Knowledge-Gap in Drone Planning via LLM-based Active Elicitation
Zeyu Fang, Beomyeol Yu, Cheng Liu +5
Human-AI joint planning in Unmanned Aerial Vehicles (UAVs) typically relies on control handover when facing environmental uncertainties, which is often inefficient and cognitively…
Knowing When to Ask: Resolving Uncertainty in Human-Robot Joint Planning via Explicit Dialogue and Implicit Intent Cues
Zeyu Fang, Yuxin Lin, Cheng Liu +6
Effective human-robot collaboration in open-world environments requires joint planning under uncertainty about the task, the environment, and the human teammate. Communication is t…
VLM-Assisted Continual learning for Visual Question Answering in Self-Driving
Yuxin Lin, Mengshi Qi, Liang Liu +1
In this paper, we propose a novel approach for solving the Visual Question Answering (VQA) task in autonomous driving by integrating Vision-Language Models (VLMs) with continual le…
RoboDriveVLM: A Novel Benchmark and Baseline towards Robust Vision-Language Models for Autonomous Driving
Dacheng Liao, Mengshi Qi, Peng Shu +4
Current Vision-Language Model (VLM)-based end-to-end autonomous driving systems often leverage large language models to generate driving decisions directly based on their understan…
HairShifter: Consistent and High-Fidelity Video Hair Transfer via Anchor-Guided Animation
Wangzheng Shi, Yinglin Zheng, Yuxin Lin +3
Hair transfer is increasingly valuable across domains such as social media, gaming, advertising, and entertainment. While significant progress has been made in single-image hair tr…