2 citations · 2 across the 2 of their papers we have counts for
1 paper · 1 filter
Shuhei Kurita, Naoki Katsura, Eri Onami
Grounding textual expressions on scene objects from first-person views is a truly demanding capability in developing agents that are aware of their surroundings and behave followin…