Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
DeepVision-103K: A Visually Diverse, Broad-Coverage, and Verifiable Mathematical Dataset for Multimodal Reasoning
Haoxiang Sun, Lizhen Xu, Bing Zhao +5
Reinforcement Learning with Verifiable Rewards (RLVR) has been shown effective in enhancing the visual reflection and reasoning capabilities of Large Multimodal Models (LMMs). Howe…
cs.LG2021
Operator Splitting for Learning to Predict Equilibria in Convex Games
Daniel McKenzie, Howard Heaton, Qiuwei Li +3
Systems of competing agents can often be modeled as games. Assuming rationality, the most likely outcomes are given by an equilibrium (e.g. a Nash equilibrium). In many practical s…