◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yuankai Qi

10 papers hereh-index 5130 citations13 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author5
  • last author5

Across the 10 of 10 papers where every author was matched, so the position is known.

fields
  • cs.SD5
  • cs.CV4
  • cs.MM1
same name
  • Yuankai Qi — 16 papers, h 32
  • Yuankai Qi — 11 papers, h 8
  • Yuankai Qi — 9 papers, h 4
  • Yuankai Qi — 3 papers, h 2
  • Yuankai Qi — 2 papers
  • Yuankai Qi — 2 papers, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedEmoDubber: Towards High Quality and Emotion Controllable Movie Dubbing

1 citations · 1 across the 10 of their papers we have counts for

collaborators
Showing cs.CVShow all

4 papers · 1 filter

cs.CV2026

Question-guided Visual Compression with Memory Feedback for Long-Term Video Understanding

Sosuke Yamao, Natsuki Miyahara, Yuankai Qi +1

In the context of long-term video understanding with large multimodal models, many frameworks have been proposed. Although transformer-based visual compressors and memory-augmented…

cs.CV2025

ProgRoCC: A Progressive Approach to Rough Crowd Counting

Shengqin Jiang, Linfei Li, Haokui Zhang +6

As the number of individuals in a crowd grows, enumeration-based techniques become increasingly infeasible and their estimates increasingly unreliable. We propose instead an estima…

cs.CV2025

Visual and Semantic Prompt Collaboration for Generalized Zero-Shot Learning

Huajie Jiang, Zhengxian Li, Xiaohan Yu +4

Generalized zero-shot learning aims to recognize both seen and unseen classes with the help of semantic information that is shared among different classes. It inevitably requires c…

cs.CV2025

Collaborative Temporal Consistency Learning for Point-supervised Natural Language Video Localization

Zhuo Tao, Liang Li, Qi Chen +5

Natural language video localization (NLVL) is a crucial task in video understanding that aims to localize the target moment in videos specified by a given language description. Rec…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.