3 papers
cs.CV2025
Advanced Object Detection and Pose Estimation with Hybrid Task Cascade and High-Resolution Networks
Yuhui Jin, Yaqiong Zhang, Zheyuan Xu +2
In the field of computer vision, 6D object detection and pose estimation are critical for applications such as robotics, augmented reality, and autonomous driving. Traditional meth…
cs.CL2025
Mitigating Knowledge Conflicts in Language Model-Driven Question Answering
Han Cao, Zhaoyang Zhang, Xiangtian Li +3
In the context of knowledge-driven seq-to-seq generation tasks, such as document-based question answering and document summarization systems, two fundamental knowledge sources play…
cs.CV2025
Generating Multimodal Images with GAN: Integrating Text, Image, and Style
Chaoyi Tan, Wenqing Zhang, Zhen Qi +3
In the field of computer vision, multimodal image generation has become a research hotspot, especially the task of integrating text, image, and style. In this study, we propose a m…