collaborators

11 papers

cs.CV2026

The 10th AI City Challenge

Zheng Tang, Shuo Wang, David C. Anastasiu +34

The 10th AI City Challenge, held with ECCV 2026, marks a decade of community benchmarking for intelligent transportation, smart cities, and physical AI. Since its 2017 start with v…

cs.CV2026

One Trajectory, One Token: Grounded Video Tokenization via Panoptic Sub-object Trajectory

Chenhao Zheng, Jieyu Zhang, Mohammadreza Salehi +5

Effective video tokenization is critical for scaling transformer models for long videos. Current approaches tokenize videos using space-time patches, leading to excessive tokens an…

cs.RO2026

Real-world Latency Analysis of Vehicular Visible Light Communication with Multiple LED Transmitters and an Event-Based Camera

Ryota Soga, Tsukasa Shimizu, Shintaro Shiba +3

Event cameras offer high temporal resolution, low latency, and wide dynamic range, making them promising receivers for visible light communication (VLC) in vehicle-to-everything (V…

cs.CV2026

InstAP: Instance-Aware Vision-Language Pre-Train for Spatial-Temporal Understanding

Ashutosh Kumar, Rajat Saini, Jingjing Pan +5

Current vision-language pre-training (VLP) paradigms excel at global scene understanding but struggle with instance-level reasoning due to global-only supervision. We introduce Ins…

cs.RO2025

Performance Evaluation of an Integrated System for Visible Light Communication and Positioning Using an Event Camera

Ryota Soga, Masataka Kobayashi, Tsukasa Shimizu +4

Event cameras, featuring high temporal resolution and high dynamic range, offer visual sensing capabilities comparable to conventional image sensors while capturing fast-moving obj…

cs.CV2025

The 9th AI City Challenge

Zheng Tang, Shuo Wang, David C. Anastasiu +25

The ninth AI City Challenge continues to advance real-world applications of computer vision and AI in transportation, industrial automation, and public safety. The 2025 edition fea…