2 papers
cs.CV2025
UniQ: Unified Decoder with Task-specific Queries for Efficient Scene Graph Generation
Xinyao Liao, Wei Wei, Dangyang Chen +1
Scene Graph Generation(SGG) is a scene understanding task that aims at identifying object entities and reasoning their relationships within a given image. In contrast to prevailing…
cs.CV2023
Detection-based Intermediate Supervision for Visual Question Answering
Yuhang Liu, Daowan Peng, Wei Wei +3
Recently, neural module networks (NMNs) have yielded ongoing success in answering compositional visual questions, especially those involving multi-hop visual and logical reasoning.…