3 papers
cs.CV2025
Improving VQA Reliability: A Dual-Assessment Approach with Self-Reflection and Cross-Model Verification
Xixian Wu, Yang Ou, Pengchao Tian +4
Vision-language models (VLMs) have demonstrated significant potential in Visual Question Answering (VQA). However, the susceptibility of VLMs to hallucinations can lead to overconf…
cs.LG2024
The Solution for the sequential task continual learning track of the 2nd Greater Bay Area International Algorithm Competition
Sishun Pan, Xixian Wu, Tingmin Li +4
This paper presents a data-free, parameter-isolation-based continual learning algorithm we developed for the sequential task continual learning track of the 2nd Greater Bay Area In…
eess.IV2024
High-Resolution Image Translation Model Based on Grayscale Redefinition
Xixian Wu, Dian Chao, Yang Yang
Image-to-image translation is a technique that focuses on transferring images from one domain to another while maintaining the essential content representations. In recent years, i…