Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models
Xiangxiang Zhang, Jingxuan Wei, Donghong Zhong +31
Existing Vision-Language Models often struggle with complex, multi-question reasoning tasks where partial correctness is crucial for effective learning. Traditional reward mechanis…
cs.AI2025
SketchAgent: Generating Structured Diagrams from Hand-Drawn Sketches
Cheng Tan, Qi Chen, Jingxuan Wei +6
Hand-drawn sketches are a natural and efficient medium for capturing and conveying ideas. Despite significant advancements in controllable natural image generation, translating fre…