3 papers
cs.CV2026
RISE: Roadside Infrastructure Sequence Understanding across 3D Tracking and Structured Vision-Language Reasoning
Yanbo Jiang, Haotian Zheng, Jiahao Wang +10
We present RISE (Roadside Infrastructure Sequence Understanding and Evaluation), a framework spanning metric 3D tracking and structured vision-language reasoning in roadside sequen…
cs.CV2026
From Scene to Object: Text-Guided Dual-Gaze Prediction
Zehong Ke, Yanbo Jiang, Jinhao Li +5
Interpretable driver attention prediction is crucial for human-like autonomous driving. However, existing datasets provide only scene-level global gaze rather than fine-grained obj…
cs.RO2026
MISTY: High-Throughput Motion Planning via Mixer-based Single-step Drifting
Yining Xing, Zehong Ke, Yiqian Tu +3
Multi-modal trajectory generation is essential for safe autonomous driving, yet existing diffusion-based planners suffer from high inference latency due to iterative neural functio…