most citedTowards Flexible Visual Relationship Segmentation

1 citations · 1 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV20241 cited

Towards Flexible Visual Relationship Segmentation

Fangrui Zhu, Jianwei Yang, Huaizu Jiang

Visual relationship understanding has been studied separately in human-object interaction(HOI) detection, scene graph generation(SGG), and referring relationships(RR) tasks. Given…

cs.CV2024

SMooDi: Stylized Motion Diffusion Model

Lei Zhong, Yiming Xie, Varun Jampani +2

We introduce a novel Stylized Motion Diffusion model, dubbed SMooDi, to generate stylized motion driven by content texts and style motion sequences. Unlike existing methods that ei…

cs.RO2024

StereoNavNet: Learning to Navigate using Stereo Cameras with Auxiliary Occupancy Voxels

Hongyu Li, Taskin Padir, Huaizu Jiang

Visual navigation has received significant attention recently. Most of the prior works focus on predicting navigation actions based on semantic features extracted from visual encod…

cs.CV2024

NeuFlow: Real-time, High-accuracy Optical Flow Estimation on Robots Using Edge Devices

Zhiyong Zhang, Huaizu Jiang, Hanumant Singh

Real-time high-accuracy optical flow estimation is a crucial component in various applications, including localization and mapping in robotics, object tracking, and activity recogn…

cs.CV2023

Pixel-Aligned Recurrent Queries for Multi-View 3D Object Detection

Yiming Xie, Huaizu Jiang, Georgia Gkioxari +1

We present PARQ - a multi-view 3D object detector with transformer and pixel-aligned recurrent queries. Unlike previous works that use learnable features or only encode 3D point po…