2 papers
cs.RO2025
TRACE: Textual Reasoning for Affordance Coordinate Extraction
Sangyun Park, Jin Kim, Yuchen Cui +1
Vision-Language Models (VLMs) struggle to translate high-level instructions into the precise spatial affordances required for robotic manipulation. While visual Chain-of-Thought (C…
cs.CV2025
Autonomous Computer Vision Development with Agentic AI
Jin Kim, Muhammad Wahi-Anwa, Sangyun Park +3
Agentic Artificial Intelligence (AI) systems leveraging Large Language Models (LLMs) exhibit significant potential for complex reasoning, planning, and tool utilization. We demonst…