◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Chao Huang

University of Rochester

24 papers hereh-index 11621 citations32 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author8
  • middle author16

Across the 24 of 24 papers where every author was matched, so the position is known.

fields
  • cs.CV19
  • cs.SD4
  • cs.CL1
affiliations
  • University of Rochester
Homepage
same name
  • Chao Huang — 44 papers, h 34
  • Chao Huang — 18 papers, h 16
  • Chao Huang — 16 papers, h 14
  • Chao Huang — 14 papers, h 30
  • Chao Huang — 13 papers, h 21
  • Chao Huang — 12 papers, h 8

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20232026
most citedVideo Understanding with Large Language Models: A Survey

8 citations · 19 across the 22 of their papers we have counts for

collaborators
Showing 2024Show all

4 papers · 1 filter

cs.CV2024

Language-Guided Joint Audio-Visual Editing via One-Shot Adaptation

Susan Liang, Chao Huang, Yapeng Tian +2

In this paper, we introduce a novel task called language-guided joint audio-visual editing. Given an audio and image pair of a sounding event, this task aims at generating new audi…

cs.CV2024

VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?

Yolo Y. Tang, Junjia Guo, Hang Hua +9

The advancement of Multimodal Large Language Models (MLLMs) has enabled significant progress in multimodal understanding, expanding their capacity to analyze video content. However…

cs.CV2024

Scaling Concept With Text-Guided Diffusion Models

Chao Huang, Susan Liang, Yunlong Tang +3

Text-guided diffusion models have revolutionized generative tasks by producing high-fidelity content from text descriptions. They have also enabled an editing paradigm where concep…

cs.SD2024

Modeling and Driving Human Body Soundfields through Acoustic Primitives

Chao Huang, Dejan Markovic, Chenliang Xu +1

While rendering and animation of photorealistic 3D human body models have matured and reached an impressive quality over the past years, modeling the spatial audio associated with…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.