dense reward learning 1failure synthesis 1reinforcement learning 1robotic manipulation 1vision-language models 1
From the 1 of 6 linked papers with an AI index.
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2024
Planner3D: LLM-enhanced graph prior meets 3D indoor scene explicit regularization
Yao Wei, Martin Renqiang Min, George Vosselman +2
Compositional 3D scene synthesis has diverse applications across a spectrum of industries such as robotics, films, and video games, as it closely mirrors the complexity of real-wor…
cs.CV2024
SOK-Bench: A Situated Video Reasoning Benchmark with Aligned Open-World Knowledge
Andong Wang, Bo Wu, Sunli Chen +5
Learning commonsense reasoning from visual contexts and scenes in real-world is a crucial step toward advanced artificial intelligence. However, existing video reasoning benchmarks…
cs.CV2024
AffordanceLLM: Grounding Affordance from Vision Language Models
Shengyi Qian, Weifeng Chen, Min Bai +3
Affordance grounding refers to the task of finding the area of an object with which one can interact. It is a fundamental but challenging task, as a successful solution requires th…