◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Takuya Hasumi

4 papers hereh-index 344 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author1

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.SD3
  • eess.AS1

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.SD2026

Aligning MusicLLM with Emotion using Instruction Tuning and Feedback-Driven Alignment

Takuya Hasumi, Welly Naptali

This paper investigates whether music large language models (MusicLLMs) can be aligned for emotion regression. While MusicLLMs have shown strong performance in music information re…

cs.SD2025

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Takuya Hasumi, Yusuke Fujita

We propose a new dataset for cinematic audio source separation (CASS) that handles non-verbal sounds. Existing CASS datasets only contain reading-style sounds as a speech stem. The…

eess.AS2025

BitTTS: Highly Compact Text-to-Speech Using 1.58-bit Quantization and Weight Indexing

Masaya Kawamura, Takuya Hasumi, Yuma Shirahata +1

This paper proposes a highly compact, lightweight text-to-speech (TTS) model for on-device applications. To reduce the model size, the proposed model introduces two techniques. Fir…

cs.SD2025

Music Tagging with Classifier Group Chains

Takuya Hasumi, Tatsuya Komatsu, Yusuke Fujita

We propose music tagging with classifier chains that model the interplay of music tags. Most conventional methods estimate multiple tags independently by treating them as multiple…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.