Publications (15)
Just ASK: Building an Architecture for Extensible Self-Service Spoken Language Understanding
Anjishnu Kumar, Arpit Gupta, Julian Chan +9
T-Mimi: A Transformer-based Mimi Decoder for Real-Time On-Phone TTS
Haibin Wu, Bach Viet Do, Naveen Suda +10
Contextualized Streaming End-to-End Speech Recognition with Trie-Based Deep Biasing and Shallow Fusion
Duc Le, Mahaveer Jain, Gil Keren +9
Converging on the eccentricity of massive black hole binaries in galactic mergers
Alessia Gualandris, Justin Read, Federica Fastidio +4
Continuous-Time Modelling of Black Hole Binary Evolution with Neural ODEs
Julian Chan, Alessia Gualandris, Payel Das
Deep Shallow Fusion for RNN-T Personalization
Duc Le, Gil Keren, Julian Chan +3
Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement
Xueyao Zhang, Xiaohui Zhang, Kainan Peng +10
Transformer in action: a comparative study of transformer-based acoustic models for large scale speech recognition applications
Yongqiang Wang, Yangyang Shi, Frank Zhang +4
Emformer: Efficient Memory Transformer Based Acoustic Model For Low Latency Streaming Speech Recognition
Yangyang Shi, Yongqiang Wang, Chunyang Wu +5
Streaming Attention-Based Models with Augmented Memory for End-to-End Speech Recognition
Ching-Feng Yeh, Yongqiang Wang, Yangyang Shi +4
Benchmarking LF-MMI, CTC and RNN-T Criteria for Streaming ASR
Xiaohui Zhang, Frank Zhang, Chunxi Liu +8
On lattice-free boosted MMI training of HMM and CTC-based full-context ASR models
Xiaohui Zhang, Vimal Manohar, David Zhang +7
Perturber-Driven Dynamics of Supermassive Black Hole Binaries in Galaxy Merger
Julian Chan, Alessia Gualandris, Walter Dehnen +1
The paper uses high‑resolution N‑body simulations of a major galaxy merger to test how massive perturbers in the host galaxy affect the eccentricity scatter of supermassive black h…
Odds are the sign is right
Brian Knaeble, Julian Chan
Dynamic Encoder Transducer: A Flexible Solution For Trading Off Accuracy For Latency
Yangyang Shi, Varun Nagaraja, Chunyang Wu +9