4 papers
RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning
Jinrui Liu, Bingyan Nie, Boyu Li +4
Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks successfully. Despite the succe…
Incomplete Multi-view Multi-label Classification via a Dual-level Contrastive Learning Framework
Bingyan Nie, Wulin Xie, Jiang Long +1
Recently, multi-view and multi-label classification have become significant domains for comprehensive data analysis and exploration. However, incompleteness both in views and label…
MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models
Wulin Xie, Yi-Fan Zhang, Chaoyou Fu +6
Existing MLLM benchmarks face significant challenges in evaluating Unified MLLMs (U-MLLMs) due to: 1) lack of standardized benchmarks for traditional tasks, leading to inconsistent…
Multi-View Factorizing and Disentangling: A Novel Framework for Incomplete Multi-View Multi-Label Classification
Wulin Xie, Lian Zhao, Jiang Long +2
Multi-view multi-label classification (MvMLC) has recently garnered significant research attention due to its wide range of real-world applications. However, incompleteness in view…