10 citations · 10 across the 3 of their papers we have counts for
3 papers
cs.CV2026
DiffBench Meets DiffAgent: End-to-End LLM-Driven Diffusion Acceleration Code Generation
Jiajun jiao, Haowei Zhu, Puyuan Yang +8
Diffusion models have achieved remarkable success in image and video generation. However, their inherently multiple step inference process imposes substantial computational overhea…
cs.AI2024
SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark
Zhenwen Liang, Kehan Guo, Gang Liu +7
The paper introduces SceMQA, a novel benchmark for scientific multimodal question answering at the college entrance level. It addresses a critical educational phase often overlooke…
cs.AI2023★ 10 cited
Elucidating STEM Concepts through Generative AI: A Multi-modal Exploration of Analogical Reasoning
Chen Cao, Zijian Ding, Gyeong-Geon Lee +3
This study explores the integration of generative artificial intelligence (AI), specifically large language models, with multi-modal analogical reasoning as an innovative approach…