◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yilun Du

15 papers hereh-index 121k citations19 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author14

Across the 14 of 15 papers where every author was matched, so the position is known.

fields
  • cs.CV9
  • cs.RO5
  • cs.AI1
same name
  • Yilun Du — 68 papers, h 48
  • Yilun Du — 23 papers, h 12
  • Yilun Du — 14 papers, h 2
  • Yilun Du — 7 papers, h 3
  • Yilun Du — 7 papers, h 4
  • Yilun Du — 6 papers, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20232026
most cited3D-VLA: A 3D Vision-Language-Action Generative World Model

14 citations · 28 across the 12 of their papers we have counts for

collaborators
Showing 2025 · cs.CVShow all

4 papers · 2 filters

cs.CV2025

MindJourney: Test-Time Scaling with World Models for Spatial Reasoning

Yuncong Yang, Jiageng Liu, Zheyuan Zhang +5

Spatial reasoning in 3D space is central to human cognition and indispensable for embodied tasks such as navigation and manipulation. However, state-of-the-art vision-language mode…

cs.CV2025

Learning 3D Persistent Embodied World Models

Siyuan Zhou, Yilun Du, Yuncong Yang +4

The ability to simulate the effects of future actions on the world is a crucial ability of intelligent embodied agents, enabling agents to anticipate the effects of their actions a…

cs.CV2025

TesserAct: Learning 4D Embodied World Models

Haoyu Zhen, Qiao Sun, Hongxin Zhang +4

This paper presents an effective approach for learning novel 4D embodied world models, which predict the dynamic evolution of 3D scenes over time in response to an embodied agent's…

cs.CV2025

Towards Understanding Camera Motions in Any Video

Zhiqiu Lin, Siyuan Cen, Daniel Jiang +12

We introduce CameraBench, a large-scale dataset and benchmark designed to assess and improve camera motion understanding. CameraBench consists of ~3,000 diverse internet videos, an…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.