◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

H. Cheng

6 papers hereh-index 4404 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author5
  • middle author1

Across the 6 of 6 papers where every author was matched, so the position is known.

fields
  • cs.CV4
  • cs.LG2
same name
  • H. Cheng — 151 papers, h 33
  • H. Cheng — 75 papers, h 49
  • H. Cheng — 40 papers, h 52
  • H. Cheng — 31 papers, h 54
  • H. Cheng — 17 papers, h 23
  • H. Cheng — 14 papers, h 34

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20232026
most citedTracking Anything with Decoupled Video Segmentation

1 citations · 1 across the 3 of their papers we have counts for

collaborators
Showing cs.CVShow all

4 papers · 1 filter

cs.CV2025

HOP: Heterogeneous Topology-based Multimodal Entanglement for Co-Speech Gesture Generation

Hongye Cheng, Tianyu Wang, Guangsi Shi +2

Co-speech gestures are crucial non-verbal cues that enhance speech clarity and expressiveness in human communication, which have attracted increasing attention in multimodal resear…

cs.CV2024

MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis

Ho Kei Cheng, Masato Ishii, Akio Hayakawa +3

We propose to synthesize high-quality and synchronized audio, given video and optional text conditions, using a novel multimodal joint training framework MMAudio. In contrast to si…

cs.CV2023★ 4 cited

Putting the Object Back into Video Object Segmentation

Ho Kei Cheng, Seoung Wug Oh, Brian Price +2

We present Cutie, a video object segmentation (VOS) network with object-level memory reading, which puts the object representation from memory back into the video object segmentati…

cs.CV2023★ 1 cited

Tracking Anything with Decoupled Video Segmentation

Ho Kei Cheng, Seoung Wug Oh, Brian Price +2

Training data for video segmentation are expensive to annotate. This impedes extensions of end-to-end algorithms to new video segmentation tasks, especially in large-vocabulary set…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.