◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Siqi Zheng

12 papers hereh-index 111.2k citations16 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author9
  • last author1

Across the 10 of 12 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.SD4
  • eess.AS4
same name
  • Siqi Zheng — 24 papers, h 12
  • Siqi Zheng — 14 papers, h 3
  • Siqi Zheng — 8 papers, h 9
  • Siqi Zheng — 2 papers, h 3
  • Siqi Zheng — 1 paper, h 46
  • Siqi Zheng — 1 paper, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20232026
most citedLauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT

18 citations · 36 across the 12 of their papers we have counts for

collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2024

OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation

Qinglin Zhang, Luyao Cheng, Chong Deng +8

Full-duplex spoken dialogue systems significantly surpass traditional turn-based dialogue systems, as they allow simultaneous bidirectional communication, closely mirroring human-h…

cs.CL2024

Skip-Layer Attention: Bridging Abstract and Detailed Dependencies in Transformers

Qian Chen, Wen Wang, Qinglin Zhang +7

The Transformer architecture has significantly advanced deep learning, particularly in natural language processing, by effectively managing long-range dependencies. However, as the…

cs.CL2024★ 4 cited

An Embarrassingly Simple Approach for LLM with Strong ASR Capacity

Ziyang Ma, Guanrou Yang, Yifan Yang +8

In this paper, we focus on solving one of the most important tasks in the field of speech processing, i.e., automatic speech recognition (ASR), with speech foundation encoders and…

cs.CL2023

Loss Masking Is Not Needed in Decoder-only Transformer for Discrete-token-based ASR

Qian Chen, Wen Wang, Qinglin Zhang +7

Recently, unified speech-text models, such as SpeechGPT, VioLA, and AudioPaLM, have achieved remarkable performance on various speech tasks. These models discretize speech signals…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.