◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Hao Shi

8 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author5

Across the 7 of 8 papers where every author was matched, so the position is known.

fields
  • cs.CV5
  • cs.CL1
  • cs.IT1
  • cs.SD1
ORCID 0000-0003-0184-2245
same name
  • Hao Shi — 16 papers, h 20
  • Hao Shi — 5 papers
  • Hao Shi — 2 papers
  • Hao Shi — 2 papers
  • Hao Shi — 1 paper, h 14
  • Hao Shi — 1 paper, h 17

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedDelivering Arbitrary-Modal Semantic Segmentation

5 citations · 8 across the 8 of their papers we have counts for

collaborators

4 papers

cs.CV2023

FishDreamer: Towards Fisheye Semantic Completion via Unified Image Outpainting and Segmentation

Hao Shi, Yu Li, Kailun Yang +7

This paper raises the new task of Fisheye Semantic Completion (FSC), where dense texture, structure, and semantics of a fisheye image are inferred even beyond the sensor field-of-v…

cs.SD2023

Time-domain Speech Enhancement Assisted by Multi-resolution Frequency Encoder and Decoder

Hao Shi, Masato Mimura, Longbiao Wang +2

Time-domain speech enhancement (SE) has recently been intensively investigated. Among recent works, DEMUCS introduces multi-resolution STFT loss to enhance performance. However, so…

cs.CV2023★ 5 cited

Delivering Arbitrary-Modal Semantic Segmentation

Jiaming Zhang, Ruiping Liu, Hao Shi +6

Multimodal fusion can make semantic segmentation more robust. However, fusing an arbitrary number of modalities remains underexplored. To delve into this problem, we create the DeL…

cs.CL2022

Language-specific Characteristic Assistance for Code-switching Speech Recognition

Tongtong Song, Qiang Xu, Meng Ge +5

Dual-encoder structure successfully utilizes two language-specific encoders (LSEs) for code-switching speech recognition. Because LSEs are initialized by two pre-trained language-s…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.