10 citations · 11 across the 4 of their papers we have counts for
4 papers
See, Hear, and Feel: Smart Sensory Fusion for Robotic Manipulation
Hao Li, Yizhi Zhang, Junzhe Zhu +7
Humans use all of their senses to accomplish different tasks in everyday activities. In contrast, existing work on robotic manipulation mostly relies on one, or occasionally two mo…
Multi-Decoder DPRNN: High Accuracy Source Counting and Separation
Junzhe Zhu, Raymond Yeh, Mark Hasegawa-Johnson
We propose an end-to-end trainable approach to single-channel speech separation with unknown number of speakers. Our approach extends the MulCat source separation backbone with add…
A Comparison Study on Infant-Parent Voice Diarization
Junzhe Zhu, Mark Hasegawa-Johnson, Nancy McElwain
We design a framework for studying prelinguistic child voicefrom 3 to 24 months based on state-of-the-art algorithms in di-arization. Our system consists of a time-invariant featur…
Identify Speakers in Cocktail Parties with End-to-End Attention
Junzhe Zhu, Mark Hasegawa-Johnson, Leda Sari
In scenarios where multiple speakers talk at the same time, it is important to be able to identify the talkers accurately. This paper presents an end-to-end system that integrates…