3 papers
cs.CV2026
RoMo: A Large-Scale, Richly Organized Dataset and Semantic Taxonomy for Human Motion Generation
Jiahao Zhang, Joseph Liu, Young-Yoon Lee +9
Success in generative modeling across language, image, and video demonstrates that large, well-curated datasets are the key driver for building capable models. 3D Human motion, how…
cs.LG2020
Iterative Compression of End-to-End ASR Model using AutoML
Abhinav Mehrotra, Łukasz Dudziak, Jinsu Yeo +9
Increasing demand for on-device Automatic Speech Recognition (ASR) systems has resulted in renewed interests in developing automatic model compression techniques. Past research hav…
eess.AS2020
Attention based on-device streaming speech recognition with large speech corpus
Kwangyoun Kim, Kyungmin Lee, Dhananjaya Gowda +10
In this paper, we present a new on-device automatic speech recognition (ASR) system based on monotonic chunk-wise attention (MoChA) models trained with large (> 10K hours) corpus.…