1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2024★ 1 cited
MMSci: A Dataset for Graduate-Level Multi-Discipline Multimodal Scientific Understanding
Zekun Li, Xianjun Yang, Kyuri Choi +11
Scientific figure interpretation is a crucial capability for AI-driven scientific assistants built on advanced Large Vision Language Models. However, current datasets and benchmark…
cs.CV2024
Automatic Layout Planning for Visually-Rich Documents with Instruction-Following Models
Wanrong Zhu, Jennifer Healey, Ruiyi Zhang +2
Recent advancements in instruction-following models have made user interactions with models more user-friendly and efficient, broadening their applicability. In graphic design, non…