◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xitong Yang

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV4
ORCID 0000-0003-4372-241X
same name
  • Xitong Yang — 4 papers, h 5
  • Xitong Yang — 4 papers, h 7
  • Xitong Yang — 3 papers, h 10
  • Xitong Yang — 3 papers, h 4
  • Xitong Yang — 1 paper, h 3
  • Xitong Yang — 1 paper, h 6

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedOpen-VCLIP: Transforming CLIP to an Open-vocabulary Video Model via Interpolated Weight Optimization

6 citations · 8 across the 4 of their papers we have counts for

collaborators
Showing cs.CVShow all

4 papers · 1 filter

cs.CV2023★ 1 cited

Towards Scalable Neural Representation for Diverse Videos

Bo He, Xitong Yang, Hanyu Wang +6

Implicit neural representations (INR) have gained increasing attention in representing 3D scenes and images, and have been recently applied to encode videos (e.g., NeRV, E-NeRV). W…

cs.CV2023★ 1 cited

MINOTAUR: Multi-task Video Grounding From Multimodal Queries

Raghav Goyal, Effrosyni Mavroudi, Xitong Yang +5

Video understanding tasks take many forms, from action detection to visual query localization and spatio-temporal grounding of sentences. These tasks differ in the type of inputs (…

cs.CV2023★ 6 cited

Open-VCLIP: Transforming CLIP to an Open-vocabulary Video Model via Interpolated Weight Optimization

Zejia Weng, Xitong Yang, Ang Li +2

Contrastive Language-Image Pretraining (CLIP) has demonstrated impressive zero-shot learning abilities for image understanding, yet limited effort has been made to investigate CLIP…

cs.CV2023

Vision Transformers Are Good Mask Auto-Labelers

Shiyi Lan, Xitong Yang, Zhiding Yu +3

We propose Mask Auto-Labeler (MAL), a high-quality Transformer-based mask auto-labeling framework for instance segmentation using only box annotations. MAL takes box-cropped images…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.