5 citations · 5 across the 1 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023
InvPT++: Inverted Pyramid Multi-Task Transformer for Visual Scene Understanding
Hanrong Ye, Dan Xu
Multi-task scene understanding aims to design models that can simultaneously predict several scene understanding tasks with one versatile model. Previous studies typically process…
cs.CV2016★ 5 cited
Searching Action Proposals via Spatial Actionness Estimation and Temporal Path Inference and Tracking
Nannan Li, Dan Xu, Zhenqiang Ying +2
In this paper, we address the problem of searching action proposals in unconstrained video clips. Our approach starts from actionness estimation on frame-level bounding boxes, and…