46 citations · 75 across the 26 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
PolyReal: A Benchmark for Real-World Polymer Science Workflows
Wanhao Liu, Weida Wang, Jiaqing Xie +12
Multimodal Large Language Models (MLLMs) excel in general domains but struggle with complex, real-world science. We posit that polymer science, an interdisciplinary field spanning…
cs.CV2025
RIV: Recursive Introspection Mask Diffusion Vision Language Model
YuQian Li, Limeng Qiao, Lin Ma
Mask Diffusion-based Vision Language Models (MDVLMs) have achieved remarkable progress in multimodal understanding tasks. However, these models are unable to correct errors in gene…
cs.CV2025★ 1 cited
AutoMat: Enabling Automated Crystal Structure Reconstruction from Microscopy via Agentic Tool Use
Yaotian Yang, Yiwen Tang, Yizhe Chen +14
Reconstructing atomistic crystal structures from a single noisy STEM projection is an ill-posed inverse problem: multiple lattices can explain similar contrast, and purely feed-for…