2 papers
cs.CV2025
YOLOA: Real-Time Affordance Detection via LLM Adapter
Yuqi Ji, Junjie Ke, Lihuo He +5
Affordance detection aims to jointly address the fundamental "what-where-how" challenge in embodied AI by understanding "what" an object is, "where" the object is located, and "how…
cs.CV2024
Multi-task Geometric Estimation of Depth and Surface Normal from Monocular 360° Images
Kun Huang, Fang-Lue Zhang, Fangfang Zhang +3
Geometric estimation is required for scene understanding and analysis in panoramic 360° images. Current methods usually predict a single feature, such as depth or surface normal.…