55 citations · 65 across the 3 of their papers we have counts for
6 papers
Object-Centric Neural Scene Rendering
Michelle Guo, Alireza Fathi, Jiajun Wu +1
We present a method for composing photorealistic scenes from captured images of objects. Our work builds upon neural radiance fields (NeRFs), which implicitly model the volumetric…
End-to-End Spoken Language Translation
Michelle Guo, Albert Haque, Prateek Verma
In this paper, we address the task of spoken language understanding. We present a method for translating spoken sentences from one language into spoken sentences in another languag…
Audio-Linguistic Embeddings for Spoken Sentences
Albert Haque, Michelle Guo, Prateek Verma +1
We propose spoken sentence embeddings which capture both acoustic and linguistic content. While existing works operate at the character, phoneme, or word level, our method learns l…
Measuring Depression Symptom Severity from Spoken Language and 3D Facial Expressions
Albert Haque, Michelle Guo, Adam S Miner +1
With more than 300 million people depressed worldwide, depression is a global problem. Due to access barriers such as social stigma, cost, and treatment availability, 60% of mental…
Privacy-Preserving Action Recognition for Smart Hospitals using Low-Resolution Depth Images
Edward Chou, Matthew Tan, Cherry Zou +4
Computer-vision hospital systems can greatly assist healthcare workers and improve medical facility treatment, but often face patient resistance due to the perceived intrusiveness…
Conditional End-to-End Audio Transforms
Albert Haque, Michelle Guo, Prateek Verma
We present an end-to-end method for transforming audio from one style to another. For the case of speech, by conditioning on speaker identities, we can train a single model to tran…