2 papers
cs.CV2025
VoQA: Visual-only Question Answering
Jianing An, Luyang Jiang, Jie Luo +2
Visual understanding requires interpreting both natural scenes and the textual information that appears within them, motivating tasks such as Visual Question Answering (VQA). Howev…
cs.LG2025
Clustering Properties of Self-Supervised Learning
Xi Weng, Jianing An, Xudong Ma +5
Self-supervised learning (SSL) methods via joint embedding architectures have proven remarkably effective at capturing semantically rich representations with strong clustering prop…