4 papers
Advanced Object Detection and Pose Estimation with Hybrid Task Cascade and High-Resolution Networks
Yuhui Jin, Yaqiong Zhang, Zheyuan Xu +2
In the field of computer vision, 6D object detection and pose estimation are critical for applications such as robotics, augmented reality, and autonomous driving. Traditional meth…
Mitigating Knowledge Conflicts in Language Model-Driven Question Answering
Han Cao, Zhaoyang Zhang, Xiangtian Li +3
In the context of knowledge-driven seq-to-seq generation tasks, such as document-based question answering and document summarization systems, two fundamental knowledge sources play…
Generating Multimodal Images with GAN: Integrating Text, Image, and Style
Chaoyi Tan, Wenqing Zhang, Zhen Qi +3
In the field of computer vision, multimodal image generation has become a research hotspot, especially the task of integrating text, image, and style. In this study, we propose a m…
YOLO-PPA based Efficient Traffic Sign Detection for Cruise Control in Autonomous Driving
Jingyu Zhang, Wenqing Zhang, Chaoyi Tan +2
It is very important to detect traffic signs efficiently and accurately in autonomous driving systems. However, the farther the distance, the smaller the traffic signs. Existing ob…