3 papers
cs.RO2026
Sparse2Act: Learning Action-Aligned Sparse 3D Representations for Cross-Domain Robot Manipulation
Yu Guo, Chang Yu, Siyu Ma +4
Explicit 3D representations are attractive for manipulation because they expose object shape, workspace geometry, and robot-object relations in metric coordinates. However, sparse…
cs.CL2025
Value-Spectrum: Quantifying Preferences of Vision-Language Models via Value Decomposition in Social Media Contexts
Jingxuan Li, Yuning Yang, Shengqi Yang +2
The recent progress in Vision-Language Models (VLMs) has broadened the scope of multimodal applications. However, evaluations often remain limited to functional tasks, neglecting a…
cs.RO2024
Triple Regression for Camera Agnostic Sim2Real Robot Grasping and Manipulation Tasks
Yuanhong Zeng, Yizhou Zhao, Ying Nian Wu
Sim2Real (Simulation to Reality) techniques have gained prominence in robotic manipulation and motion planning due to their ability to enhance success rates by enabling agents to t…