3 citations · 3 across the 4 of their papers we have counts for
8 papers
Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation
Sung June Kim, Sangpil Kim, Honglak Lee
On-policy exploration is a crucial component for training robust Vision-Language Navigation agents, as it exposes the policy to a broader state distribution. However, such explorat…
SECOND-Grasp: Semantic Contact-guided Dexterous Grasping
Han Yi Shin, Heeju Ko, Jaewon Mun +6
Achieving reliable robotic manipulation, such as dexterous grasping, requires a synergy between physically stable interactions and semantic task guidance, yet these objectives are…
View Selection for 3D Captioning via Diffusion Ranking
Tiange Luo, Justin Johnson, Honglak Lee
Scalable annotation approaches are crucial for constructing extensive 3D-text datasets, facilitating a broader range of applications. However, existing methods sometimes lead to th…
A Definition of AGI
Dan Hendrycks, Dawn Song, Christian Szegedy +30
The lack of a concrete definition for Artificial General Intelligence (AGI) obscures the gap between today's specialized AI and human-level cognition. This paper introduces a quant…
Visual Test-time Scaling for GUI Agent Grounding
Tiange Luo, Lajanugen Logeswaran, Justin Johnson +1
We introduce RegionFocus, a visual test-time scaling approach for Vision Language Model Agents. Understanding webpages is challenging due to the visual complexity of GUI images and…
Active Test-time Vision-Language Navigation
Heeju Ko, Sungjune Kim, Gyeongrok Oh +5
Vision-Language Navigation (VLN) policies trained on offline datasets often exhibit degraded task performance when deployed in unfamiliar navigation environments at test time, wher…