Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Multi-Branch Policy Optimization for Multimodal Large Language Models
Shuai Lyu, Yuning Gong, Ruiling Gao +7
Group-based reinforcement learning methods for multimodal large language models typically rely on trajectory-level credit assignment that applies a single advantage to all tokens i…
cs.CV2024
PointCloud-Text Matching: Benchmark Datasets and a Baseline
Yanglin Feng, Yang Qin, Dezhong Peng +3
In this paper, we present and study a new instance-level retrieval task: PointCloud-Text Matching (PTM), which aims to identify the exact cross-modal instance that matches a given…