◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

K. Koishida

13 papers hereh-index 202.1k citations70 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3
  • last author9

Across the 12 of 13 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.AI3
  • cs.SD3
  • cs.LG2
  • cs.CV1

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.SDShow all

3 papers · 1 filter

cs.SD2024

Zero-Shot Text-to-Speech from Continuous Text Streams

Trung Dang, David Aponte, Dung Tran +2

Existing zero-shot text-to-speech (TTS) systems are typically designed to process complete sentences and are constrained by the maximum duration for which they have been trained. H…

cs.SD2024

ConsistencyTTA: Accelerating Diffusion-Based Text-to-Audio Generation with Consistency Distillation

Yatong Bai, Trung Dang, Dung Tran +2

Diffusion models are instrumental in text-to-audio (TTA) generation. Unfortunately, they suffer from slow inference due to an excessive number of queries to the underlying denoisin…

cs.SD2024

LiveSpeech: Low-Latency Zero-shot Text-to-Speech via Autoregressive Modeling of Audio Discrete Codes

Trung Dang, David Aponte, Dung Tran +1

Prior works have demonstrated zero-shot text-to-speech by using a generative language model on audio tokens obtained via a neural audio codec. It is still challenging, however, to…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.