papers

Publications (21)

cs.CV2019

Real time backbone for semantic segmentation

Zhengeng Yang, Hongshan Yu, Qiang Fu +4

The rapid development of autonomous driving in recent years presents lots of challenges for scene understanding. As an essential step towards scene understanding, semantic segmenta…

cs.CV2026

Symmetry-Aware 9D Pose Estimation with Sim(3)-Consistent Feature and Spherical Inception Convolution

Panfei Cheng, Hongshan Yu, Wenrui Chen +3

Object pose estimation is a fundamental problem for an agent system to perceive or manipulate objects in images or videos. However, current instance-level methods struggle with gen…

cs.CV2026

Boosting Infrared Small Target Detection via Logit-Domain Contrast and Adaptive Shape Refinement

Handong Zeng, Zhengeng Yang, Shuai Zhang +2

Infrared small target detection (IRSTD) remains challenging due to tiny target size, low signal-to-noise ratio, severe foreground-background imbalance, and blurred boundaries in co…

cs.CV2024

OST: Refining Text Knowledge with Optimal Spatio-Temporal Descriptor for General Video Recognition

Tongjia Chen, Hongshan Yu, Zhengeng Yang +3

Due to the resource-intensive nature of training vision-language models on expansive video data, a majority of studies have centered on adapting pre-trained image-language models t…

cs.CV2023

DANet: Density Adaptive Convolutional Network with Interactive Attention for 3D Point Clouds

Yong He, Hongshan Yu, Zhengeng Yang +3

Local features and contextual dependencies are crucial for 3D point cloud analysis. Many works have been devoted to designing better local convolutional kernels that exploit the co…

cs.CV2026

Efficient Point Cloud Processing with High-Dimensional Positional Encoding and Non-Local MLPs

Yanmei Zou, Hongshan Yu, Yaonan Wang +4

Multi-Layer Perceptron (MLP) models are the foundation of contemporary point cloud processing. However, their complex network architectures obscure the source of their strength and…