1 paper
Homagni Saha, Fateme Fotouhif, Qisai Liu +1
In this paper we propose a new framework - MoViLan (Modular Vision and Language) for execution of visually grounded natural language instructions for day to day indoor household ta…