◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

David Harwath

25 papers hereh-index 181.3k citations62 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author10
  • last author14

Across the 24 of 25 papers where every author was matched, so the position is known.

fields
  • eess.AS13
  • cs.CL7
  • cs.CV3
  • cs.SD2

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.CVShow all

3 papers · 1 filter

cs.CV2025

FacEDiT: Unified Talking Face Editing and Generation via Facial Motion Infilling

Kim Sung-Bin, Joohyun Chang, David Harwath +1

Talking face editing and face generation have often been studied as distinct problems. In this work, we propose viewing both not as separate tasks but as subtasks of a unifying for…

cs.CV2025

VoiceCraft-Dub: Automated Video Dubbing with Neural Codec Language Models

Kim Sung-Bin, Jeongsoo Choi, Puyuan Peng +3

We present VoiceCraft-Dub, a novel approach for automated video dubbing that synthesizes high-quality speech from text and facial cues. This task has broad applications in filmmaki…

cs.CV2024

Action2Sound: Ambient-Aware Generation of Action Sounds from Egocentric Videos

Changan Chen, Puyuan Peng, Ami Baid +4

Generating realistic audio for human actions is important for many applications, such as creating sound effects for films or virtual reality games. Existing approaches implicitly a…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.