activity
20182020
most citedObject-Centric Neural Scene Rendering

55 citations · 65 across the 3 of their papers we have counts for

collaborators

6 papers

cs.CV202055 cited

Object-Centric Neural Scene Rendering

Michelle Guo, Alireza Fathi, Jiajun Wu +1

We present a method for composing photorealistic scenes from captured images of objects. Our work builds upon neural radiance fields (NeRFs), which implicitly model the volumetric…

cs.CL20195 cited

End-to-End Spoken Language Translation

Michelle Guo, Albert Haque, Prateek Verma

In this paper, we address the task of spoken language understanding. We present a method for translating spoken sentences from one language into spoken sentences in another languag…

cs.SD20195 cited

Audio-Linguistic Embeddings for Spoken Sentences

Albert Haque, Michelle Guo, Prateek Verma +1

We propose spoken sentence embeddings which capture both acoustic and linguistic content. While existing works operate at the character, phoneme, or word level, our method learns l…

cs.CV2018

Measuring Depression Symptom Severity from Spoken Language and 3D Facial Expressions

Albert Haque, Michelle Guo, Adam S Miner +1

With more than 300 million people depressed worldwide, depression is a global problem. Due to access barriers such as social stigma, cost, and treatment availability, 60% of mental…

cs.CV2018

Privacy-Preserving Action Recognition for Smart Hospitals using Low-Resolution Depth Images

Edward Chou, Matthew Tan, Cherry Zou +4

Computer-vision hospital systems can greatly assist healthcare workers and improve medical facility treatment, but often face patient resistance due to the perceived intrusiveness…

cs.SD2018

Conditional End-to-End Audio Transforms

Albert Haque, Michelle Guo, Prateek Verma

We present an end-to-end method for transforming audio from one style to another. For the case of speech, by conditioning on speaker identities, we can train a single model to tran…