4 papers
PhysInOne: Visual Physics Learning and Reasoning in One Suite
Siyuan Zhou, Hejun Wang, Hu Cheng +36
We present PhysInOne, a large-scale synthetic dataset addressing the critical scarcity of physically-grounded training data for AI systems. Unlike existing datasets limited to mere…
IAG: Input-aware Backdoor Attack on VLM-based Visual Grounding
Junxian Li, Beining Xu, Simin Chen +4
Recent advances in vision-language models (VLMs) have significantly enhanced the visual grounding task, which involves locating objects in an image based on natural language querie…
Chem-R: Learning to Reason as a Chemist
Weida Wang, Benteng Chen, Di Zhang +14
Although large language models (LLMs) have significant potential to advance chemical discovery, current LLMs lack core chemical knowledge, produce unreliable reasoning trajectories…
Mol-R1: Towards Explicit Long-CoT Reasoning in Molecule Discovery
Jiatong Li, Weida Wang, Qinggang Zhang +6
Large language models (LLMs), especially Explicit Long Chain-of-Thought (CoT) reasoning models like DeepSeek-R1 and QWQ, have demonstrated powerful reasoning capabilities, achievin…