Publications (21)
Real time backbone for semantic segmentation
Zhengeng Yang, Hongshan Yu, Qiang Fu +4
The rapid development of autonomous driving in recent years presents lots of challenges for scene understanding. As an essential step towards scene understanding, semantic segmenta…
Symmetry-Aware 9D Pose Estimation with Sim(3)-Consistent Feature and Spherical Inception Convolution
Panfei Cheng, Hongshan Yu, Wenrui Chen +3
Object pose estimation is a fundamental problem for an agent system to perceive or manipulate objects in images or videos. However, current instance-level methods struggle with gen…
Boosting Infrared Small Target Detection via Logit-Domain Contrast and Adaptive Shape Refinement
Handong Zeng, Zhengeng Yang, Shuai Zhang +2
Infrared small target detection (IRSTD) remains challenging due to tiny target size, low signal-to-noise ratio, severe foreground-background imbalance, and blurred boundaries in co…
OST: Refining Text Knowledge with Optimal Spatio-Temporal Descriptor for General Video Recognition
Tongjia Chen, Hongshan Yu, Zhengeng Yang +3
Due to the resource-intensive nature of training vision-language models on expansive video data, a majority of studies have centered on adapting pre-trained image-language models t…
DANet: Density Adaptive Convolutional Network with Interactive Attention for 3D Point Clouds
Yong He, Hongshan Yu, Zhengeng Yang +3
Local features and contextual dependencies are crucial for 3D point cloud analysis. Many works have been devoted to designing better local convolutional kernels that exploit the co…
Efficient Point Cloud Processing with High-Dimensional Positional Encoding and Non-Local MLPs
Yanmei Zou, Hongshan Yu, Yaonan Wang +4
Multi-Layer Perceptron (MLP) models are the foundation of contemporary point cloud processing. However, their complex network architectures obscure the source of their strength and…