◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xitong Yang

4 papers hereh-index 5203 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author3

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV4
same name
  • Xitong Yang — 6 papers
  • Xitong Yang — 1 paper, h 3
  • Xitong Yang — 1 paper, h 6
  • Xitong Yang — 1 paper
  • Xitong Yang — 1 paper, h 3
  • Xitong Yang — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20182020
most citedSTEP: Spatio-Temporal Progressive Learning for Video Action Detection

4 citations · 8 across the 2 of their papers we have counts for

collaborators

4 papers

cs.CV2020

A Generic Visualization Approach for Convolutional Neural Networks

Ahmed Taha, Xitong Yang, Abhinav Shrivastava +1

Retrieval networks are essential for searching and indexing. Compared to classification networks, attention visualization for retrieval networks is hardly studied. We formulate att…

cs.CV2019★ 4 cited

STEP: Spatio-Temporal Progressive Learning for Video Action Detection

Xitong Yang, Xiaodong Yang, Ming-Yu Liu +3

In this paper, we propose Spatio-TEmporal Progressive (STEP) action detector---a progressive learning framework for spatio-temporal action detection in videos. Starting from a hand…

cs.CV2019★ 4 cited

Exploring Uncertainty in Conditional Multi-Modal Retrieval Systems

Ahmed Taha, Yi-Ting Chen, Xitong Yang +2

We cast visual retrieval as a regression problem by posing triplet loss as a regression loss. This enables epistemic uncertainty estimation using dropout as a Bayesian approximatio…

cs.CV2018

Two Stream Self-Supervised Learning for Action Recognition

Ahmed Taha, Moustafa Meshry, Xitong Yang +2

We present a self-supervised approach using spatio-temporal signals between video frames for action recognition. A two-stream architecture is leveraged to tangle spatial and tempor…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.