2 papers
cs.CL2025
SentiMM: A Multimodal Multi-Agent Framework for Sentiment Analysis in Social Media
Xilai Xu, Zilin Zhao, Chengye Song +4
With the increasing prevalence of multimodal content on social media, sentiment analysis faces significant challenges in effectively processing heterogeneous data and recognizing m…
cs.RO2025
UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms
Xueyang Guo, Hongwei Hu, Chengye Song +5
Open-vocabulary, task-oriented grasping of specific functional parts, particularly with dual arms, remains a key challenge, as current Vision-Language Models (VLMs), while enhancin…