◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Mike Zheng Shou

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author2

Across the 2 of 3 papers where every author was matched, so the position is known.

fields
  • cs.CV3
ORCID 0000-0002-7681-2166
same name
  • Mike Zheng Shou — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20212024
most citedAVA-AVD: Audio-Visual Speaker Diarization in the Wild

33 citations · 36 across the 3 of their papers we have counts for

collaborators

3 papers

cs.CV2024★ 2 cited

Faster Diffusion via Temporal Attention Decomposition

Haozhe Liu, Wentian Zhang, Jinheng Xie +6

We explore the role of attention mechanism during inference in text-conditional diffusion models. Empirical observations suggest that cross-attention outputs converge to a fixed po…

cs.CV2023★ 1 cited

ViT-Lens: Towards Omni-modal Representations

Weixian Lei, Yixiao Ge, Kun Yi +6

Aiming to advance AI agents, large foundation models significantly improve reasoning and instruction execution, yet the current focus on vision and language neglects the potential…

cs.CV2021★ 33 cited

AVA-AVD: Audio-Visual Speaker Diarization in the Wild

Eric Zhongcong Xu, Zeyang Song, Satoshi Tsutsui +3

Audio-visual speaker diarization aims at detecting "who spoke when" using both auditory and visual signals. Existing audio-visual diarization datasets are mainly focused on indoor…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.