inference-time refinement 1multimodal reasoning 1reinforcement learning 1self-verification 1vision-language models 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.AI2026
SVR-R1: Bootstrapping Multi-modal Reasoning with Self-verification in Reinforcement Learning
Mingyuan Wu, Jingcheng Yang, Shengyi Qian +11
The paper introduces SVR-R1, a reinforcement learning framework that lets a multimodal model generate an answer and then self‑verify it with a binary verdict, allowing a second‑cha…
cs.CV2024
Unified Framework for Open-World Compositional Zero-shot Learning
Hirunima Jayasekara, Khoi Pham, Nirat Saini +1
Open-World Compositional Zero-Shot Learning (OW-CZSL) addresses the challenge of recognizing novel compositions of known primitives and entities. Even though prior works utilize la…