3 citations · 7 across the 9 of their papers we have counts for
6 papers · 1 filter
SketchAgent: Generating Structured Diagrams from Hand-Drawn Sketches
Cheng Tan, Qi Chen, Jingxuan Wei +6
Hand-drawn sketches are a natural and efficient medium for capturing and conveying ideas. Despite significant advancements in controllable natural image generation, translating fre…
UniIF: Unified Molecule Inverse Folding
Zhangyang Gao, Jue Wang, Cheng Tan +5
Molecule inverse folding has been a long-standing challenge in chemistry and biology, with the potential to revolutionize drug discovery and material science. Despite specified mod…
Progressive Multi-Modality Learning for Inverse Protein Folding
Jiangbin Zheng, Stan Z. Li
While deep generative models show promise for learning inverse protein folding directly from data, the lack of publicly available structure-sequence pairings limits their generaliz…
Boosting the Power of Small Multimodal Reasoning Models to Match Larger Models with Self-Consistency Training
Cheng Tan, Jingxuan Wei, Zhangyang Gao +5
Multimodal reasoning is a challenging task that requires models to reason across multiple modalities to answer questions. Existing approaches have made progress by incorporating la…
Enhancing Human-like Multi-Modal Reasoning: A New Challenging Dataset and Comprehensive Framework
Jingxuan Wei, Cheng Tan, Zhangyang Gao +5
Multimodal reasoning is a critical component in the pursuit of artificial intelligence systems that exhibit human-like intelligence, especially when tackling complex tasks. While t…
CUP: Critic-Guided Policy Reuse
Jin Zhang, Siyuan Li, Chongjie Zhang
The ability to reuse previous policies is an important aspect of human intelligence. To achieve efficient policy reuse, a Deep Reinforcement Learning (DRL) agent needs to decide wh…