activity
20162022
most citedVisualEchoes: Spatial Image Representation Learning through Echolocation

10 citations · 16 across the 4 of their papers we have counts for

collaborators

14 papers

cs.CV2022

Zero Experience Required: Plug & Play Modular Transfer Learning for Semantic Visual Navigation

Ziad Al-Halah, Santhosh K. Ramakrishnan, Kristen Grauman

In reinforcement learning for visual navigation, it is common to develop a model for each new task, and train that model from scratch with task-specific interactions in 3D environm…

cs.CV2021

Move2Hear: Active Audio-Visual Source Separation

Sagnik Majumder, Ziad Al-Halah, Kristen Grauman

We introduce the active audio-visual source separation problem, where an agent must move intelligently in order to better isolate the sounds coming from an object of interest in it…

cs.CV20215 cited

Environment Predictive Coding for Embodied Agents

Santhosh K. Ramakrishnan, Tushar Nagarajan, Ziad Al-Halah +1

We introduce environment predictive coding, a self-supervised approach to learn environment-level representations for embodied agents. In contrast to prior work on self-supervised…

cs.CV20201 cited

Semantic Audio-Visual Navigation

Changan Chen, Ziad Al-Halah, Kristen Grauman

Recent work on audio-visual navigation assumes a constantly-sounding target and restricts the role of audio to signaling the target's position. We introduce semantic audio-visual n…

cs.CV2020

Modeling Fashion Influence from Photos

Ziad Al-Halah, Kristen Grauman

The evolution of clothing styles and their migration across the world is intriguing, yet difficult to describe quantitatively. We propose to discover and quantify fashion influence…

cs.CV2020

Occupancy Anticipation for Efficient Exploration and Navigation

Santhosh K. Ramakrishnan, Ziad Al-Halah, Kristen Grauman

State-of-the-art navigation methods leverage a spatial memory to generalize to new environments, but their occupancy maps are limited to capturing the geometric structures directly…