Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding
Qize Yu, Jiadi You, Yuran Wang +10
Vision-Language-Action (VLA) models leverage the rich world knowledge of pretrained vision-language models (VLMs) to enable instruction-following robotic manipulation. However, the…
cs.RO2026
Affordance Agent Harness: Verification-Gated Skill Orchestration
Haojian Huang, Jiahao Shi, Yinchuan Li +1
Affordance grounding requires identifying where and how an agent should interact in open-world scenes, where actionable regions are often small, occluded, reflective, and visually…