2 citations · 5 across the 9 of their papers we have counts for
9 papers
Construction and Analysis of Impression Caption Dataset for Environmental Sounds
Yuki Okamoto, Ryotaro Nagase, Minami Okamoto +4
Some datasets with the described content and order of occurrence of sounds have been released for conversion between environmental sound and text. However, there are very few texts…
RISC: A Corpus for Shout Type Classification and Shout Intensity Prediction
Takahiro Fukumori, Taito Ishida, Yoichi Yamashita
The detection of shouted speech is crucial in audio surveillance and monitoring. Although it is desirable for a security system to be able to identify emergencies, existing corpora…
Environmental sound synthesis from vocal imitations and sound event labels
Yuki Okamoto, Keisuke Imoto, Shinnosuke Takamichi +3
One way of expressing an environmental sound is using vocal imitations, which involve the process of replicating or mimicking the rhythm and pitch of sounds by voice. We can effect…
How Should We Evaluate Synthesized Environmental Sounds
Yuki Okamoto, Keisuke Imoto, Shinnosuke Takamichi +2
Although several methods of environmental sound synthesis have been proposed, there has been no discussion on how synthesized environmental sounds should be evaluated. Only either…
Sound Event Detection Guided by Semantic Contexts of Scenes
Noriyuki Tonami, Keisuke Imoto, Ryotaro Nagase +3
Some studies have revealed that contexts of scenes (e.g., "home," "office," and "cooking") are advantageous for sound event detection (SED). Mobile devices and sensing technologies…
Sound Event Detection Based on Curriculum Learning Considering Learning Difficulty of Events
Noriyuki Tonami, Keisuke Imoto, Yuki Okamoto +2
In conventional sound event detection (SED) models, two types of events, namely, those that are present and those that do not occur in an acoustic scene, are regarded as the same t…