20 citations · 40 across the 4 of their papers we have counts for
11 papers
Anomalous Sound Detection Using a Binary Classification Model and Class Centroids
Ibuki Kuroyanagi, Tomoki Hayashi, Kazuya Takeda +1
An anomalous sound detection system to detect unknown anomalous sounds usually needs to be built using only normal sound data. Moreover, it is desirable to improve the system by ef…
Road Scene Graph: A Semantic Graph-Based Scene Representation Dataset for Intelligent Vehicles
Yafu Tian, Alexander Carballo, Ruifeng Li +1
Rich semantic information extraction plays a vital role on next-generation intelligent vehicles. Currently there is great amount of research focusing on fundamental applications su…
Characterization of Multiple 3D LiDARs for Localization and Mapping using Normal Distributions Transform
Alexander Carballo, Abraham Monrroy, David Wong +6
In this work, we present a detailed comparison of ten different 3D LiDAR sensors, covering a range of manufacturers, models, and laser configurations, for the tasks of mapping and…
LIBRE: The Multiple 3D LiDAR Dataset
Alexander Carballo, Jacob Lambert, Abraham Monrroy-Cano +6
In this work, we present LIBRE: LiDAR Benchmarking and Reference, a first-of-its-kind dataset featuring 10 different LiDAR sensors, covering a range of manufacturers, models, and l…
End-to-End Automatic Speech Recognition Integrated With CTC-Based Voice Activity Detection
Takenori Yoshimura, Tomoki Hayashi, Kazuya Takeda +1
This paper integrates a voice activity detection (VAD) function with end-to-end automatic speech recognition toward an online speech interface and transcribing very long audio reco…
ESPnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to-Speech Toolkit
Tomoki Hayashi, Ryuichi Yamamoto, Katsuki Inoue +6
This paper introduces a new end-to-end text-to-speech (E2E-TTS) toolkit named ESPnet-TTS, which is an extension of the open-source speech processing toolkit ESPnet. The toolkit sup…