activity
20202024
most citedRWCP-SSD-Onomatopoeia: Onomatopoeic Word Dataset for Environmental Sound Synthesis

2 citations · 5 across the 9 of their papers we have counts for

collaborators

9 papers

cs.SD2024

Construction and Analysis of Impression Caption Dataset for Environmental Sounds

Yuki Okamoto, Ryotaro Nagase, Minami Okamoto +4

Some datasets with the described content and order of occurrence of sounds have been released for conversion between environmental sound and text. However, there are very few texts…

cs.SD2023

RISC: A Corpus for Shout Type Classification and Shout Intensity Prediction

Takahiro Fukumori, Taito Ishida, Yoichi Yamashita

The detection of shouted speech is crucial in audio surveillance and monitoring. Although it is desirable for a security system to be able to identify emergencies, existing corpora…

cs.SD2023

Environmental sound synthesis from vocal imitations and sound event labels

Yuki Okamoto, Keisuke Imoto, Shinnosuke Takamichi +3

One way of expressing an environmental sound is using vocal imitations, which involve the process of replicating or mimicking the rhythm and pitch of sounds by voice. We can effect…

cs.SD2022

How Should We Evaluate Synthesized Environmental Sounds

Yuki Okamoto, Keisuke Imoto, Shinnosuke Takamichi +2

Although several methods of environmental sound synthesis have been proposed, there has been no discussion on how synthesized environmental sounds should be evaluated. Only either…

cs.SD2021

Sound Event Detection Guided by Semantic Contexts of Scenes

Noriyuki Tonami, Keisuke Imoto, Ryotaro Nagase +3

Some studies have revealed that contexts of scenes (e.g., "home," "office," and "cooking") are advantageous for sound event detection (SED). Mobile devices and sensing technologies…

cs.SD2021

Sound Event Detection Based on Curriculum Learning Considering Learning Difficulty of Events

Noriyuki Tonami, Keisuke Imoto, Yuki Okamoto +2

In conventional sound event detection (SED) models, two types of events, namely, those that are present and those that do not occur in an acoustic scene, are regarded as the same t…