2 papers
cs.CL2026
ProcessThinker: Enhancing Multi-modal Large Language Models Reasoning via Rollout-based Process Reward
Jingpei Wu, Xiao Han, Weixiang Shen +3
Visual question answering increasingly requires multi-step reasoning. Recent post-training with reinforcement learning under verifiable rewards (RLVR) and Group Relative Policy Opt…
cs.AI2025
Wiki-TabNER: Integrating Named Entity Recognition into Wikipedia Tables
Aneta Koleva, Martin Ringsquandl, Ahmed Hatem +2
Interest in solving table interpretation tasks has grown over the years, yet it still relies on existing datasets that may be overly simplified. This is potentially reducing the ef…