3 papers
cs.AI2024
COMET: "Cone of experience" enhanced large multimodal model for mathematical problem generation
Sannyuya Liu, Jintian Feng, Zongkai Yang +4
The automatic generation of high-quality mathematical problems is practically valuable in many educational scenarios. Large multimodal model provides a novel technical approach for…
cs.CV2023
Triple Correlations-Guided Label Supplementation for Unbiased Video Scene Graph Generation
Wenqing Wang, Kaifeng Gao, Yawei Luo +5
Video-based scene graph generation (VidSGG) is an approach that aims to represent video content in a dynamic graph by identifying visual entities and their relationships. Due to th…
cs.CV2023
Taking A Closer Look at Visual Relation: Unbiased Video Scene Graph Generation with Decoupled Label Learning
Wenqing Wang, Yawei Luo, Zhiqing Chen +4
Current video-based scene graph generation (VidSGG) methods have been found to perform poorly on predicting predicates that are less represented due to the inherent biased distribu…