1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2026
SVCBench: A Streaming Video Counting Benchmark for Spatial-Temporal State Maintenance
Pengyiang Liu, Zhongyue Shi, Hongye Hao +7
Video understanding requires models to continuously track and update world state during playback. Although existing benchmarks have advanced video understanding evaluation across m…
cs.CV2021★ 1 cited
Structure Information is the Key: Self-Attention RoI Feature Extractor in 3D Object Detection
Diankun Zhang, Zhijie Zheng, Xueting Bi +1
Unlike 2D object detection where all RoI features come from grid pixels, the RoI feature extraction of 3D point cloud object detection is more diverse. In this paper, we first comp…