From the 1 of 22 linked papers with an AI index.
22 papers
Parse, Search, and Confirmation: Training-Free Aerial Vision-and-Dialog Navigation with Chain-of-Thought Reasoning and Structured Spatial Memory
Yu Qi, Hongyu Li, Shaofei Huang +6
The paper introduces PSC-AVDN, a training‑free framework for high‑altitude UAV navigation that parses dialog instructions, searches with chain‑of‑thought reasoning, and confirms ta…
SoccerNet 2026 Challenges Results
Anthony Cioppa, Silvio Giancola, Håkan Ardö +102
The SoccerNet 2026 Challenges constitute the sixth annual edition of the SoccerNet open benchmarking effort, dedicated to advancing computer vision research in sports video underst…
Robust Trajectory Distillation: Hybrid Reweighting Meets Teacher-Inspired Targets
Kaifeng Chen, Lechao Cheng, Jiyang Li +6
Dataset distillation (DD) condenses large corpora into compact, information-rich subsets for efficient training and reuse. However, under noisy supervision, DD risks condensing cor…
CORE: Conflict-Oriented Reasoning for General Multimodal Manipulation Detection
Jinjie Shen, Yaxiong Wang, Yujiao Wu +5
The rapid rise of generative AI has made multimodal fake news increasingly realistic and pervasive, posing severe threats to public trust and social stability. Existing detection m…
REVEAL: Reference-Grounded Reasoning for Multimodal Manipulation Detection
Jun Zhou, Bingwen Hu, Yaxiong Wang +4
Multimodal manipulation detection aims to simultaneously identify forged image--text pairs and localize tampered regions, yet existing methods typically rely on memorizing isolated…
OmniVL-Guard Pro: A Tool-Augmented Agent for Omnibus Vision-Language Forensics
Jinjie Shen, Zheng Huang, Yuchen Zhang +7
Existing vision-language forgery detection and grounding methods operate under a closed-world paradigm, assuming verification can be completed by the model alone. However, self-con…