◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Qingyu Luo

4 papers hereh-index 111 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author3

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.SD2
  • cs.CV1
  • eess.AS1

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

eess.AS2026

Visually-Guided Spatial Audio Generation for 360∘ In-the-Wild Speech Scenes

Qingyu Luo, Peng Zhang, Wenwu Wang +1

Spatial audio is a key component of immersive 360∘ media, yet high-quality spatial capture remains limited in real-world speech-dominant scenes. We study visually guided Fir…

cs.SD2026

Grammar-Guided Hierarchical Parsing for Long-form Audio Activity Recognition

Peng Zhang, Qingyu Luo, Philip J. B. Jackson +1

Long-form audio exhibits an inherent hierarchy: fine-grained events form sub-activities, which in turn constitute higher-level activities. Prior work often models these levels sepa…

cs.SD2026

Hierarchical Activity Recognition and Captioning from Long-Form Audio

Peng Zhang, Qingyu Luo, Philip J. B. Jackson +1

Complex activities in real-world audio unfold over extended durations and exhibit hierarchical structure, yet most prior work focuses on short clips and isolated events. To bridge…

cs.CV2025

Teacher-Guided Pseudo Supervision and Cross-Modal Alignment for Audio-Visual Video Parsing

Yaru Chen, Ruohao Guo, Liting Gao +4

Weakly-supervised audio-visual video parsing (AVVP) seeks to detect audible, visible, and audio-visual events without temporal annotations. Previous work has emphasized refining gl…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.