collaborators

5 papers

cs.CV2026

SciFigAlign: Scoring Scientific Figures by Fine-tuned Alignment of Visuals with Manuscript Evidence

Chuanzhi Xu, Zihan Deng, Huiqi Liang +4

Scientific figure assessment in peer review differs fundamentally from general image quality evaluation: a figure must be visually legible, faithfully support the manuscript's clai…

cs.CV2026

SciFigQual-Bench: A Benchmark for Scientific Figure Quality Assessment with Full-Manuscript Context

Zihan Deng, Chuanzhi Xu, Huiqi Liang +3

Scientific images are the core elements of presenting experimental conclusions, elaborating system architecture, and supporting comparative arguments in scientific papers. However,…

cs.GR2026

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches

Chuanzhi Xu, Huiqi Liang, Bang Shi +7

Long video generation requires high-fidelity synthesis, coherent narrative structure, and user control over extended time spans. Existing text-to-video methods often rely on a sing…

cs.CV2026

CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation

Zelin Zhang, Kedi Li, Huiqi Liang +3

Multimodal semantic segmentation has shown great potential in leveraging complementary information across diverse sensing modalities. However, existing approaches often rely on car…

cs.CV2026

RxnBench: A Multimodal Benchmark for Evaluating Large Language Models on Chemical Reaction Understanding from Scientific Literature

Hanzheng Li, Xi Fang, Yixuan Li +9

The integration of Multimodal Large Language Models (MLLMs) into chemistry promises to revolutionize scientific discovery, yet their ability to comprehend the dense, graphical lang…