◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xu Cao

7 papers hereh-index 324 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author5
  • last author2

Across the 7 of 7 papers where every author was matched, so the position is known.

fields
  • cs.CV5
  • cs.AI1
  • cs.RO1
same name
  • Xu Cao — 6 papers, h 3
  • Xu Cao — 5 papers, h 2
  • Xu Cao — 5 papers, h 11
  • Xu Cao — 5 papers, h 4
  • Xu Cao — 4 papers, h 1
  • Xu Cao — 4 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2026

ChronoVision: Temporal Reasoning via Latent State Reconstruction

Yifan Shen, Jian Xu, Boyi Li +6

Multimodal large language models excel at passive perception but struggle with complex visual cognitive tasks requiring multi-step temporal reasoning. This degradation largely stem…

cs.CV2026

Decoding Children's Gait Behavior

Yifan Shen, Boyi Li, Meihuan Huang +12

We introduce a new problem domain for human action recognition: the fine-grained analysis of children's gait behaviors from standard RGB video. We specifically target the ambulator…

cs.CV2026

The 1st AI Children Challenge

Boyi Li, Yifan Shen, Houze Yang +7

The First AI Children Challenge aims to advance real-world applications of computer vision and AI in child healthcare, child education, and pediatrics. The 2026 CV4CHL edition feat…

cs.CV2026

EgoForge: Goal-Directed Egocentric World Simulator

Yifan Shen, Jiateng Liu, Xinzhuo Li +9

Generative world models have shown promise for simulating dynamic environments, yet egocentric video remains challenging due to rapid viewpoint changes, frequent hand-object intera…

cs.CV2026

Toward Cognitive Supersensing in Multimodal Large Language Model

Boyi Li, Yifan Shen, Yuanzhe Liu +12

Multimodal Large Language Models (MLLMs) have achieved remarkable success in open-vocabulary perceptual tasks, yet their ability to solve complex cognitive problems remains limited…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.