◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yong Dai

13 papers hereh-index 356 citations13 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author12

Across the 12 of 13 papers where every author was matched, so the position is known.

fields
  • cs.CV4
  • cs.AI3
  • cs.RO3
  • cs.CL2
  • cs.LG1
same name
  • Yong Dai — 7 papers, h 8
  • Yong Dai — 6 papers, h 3
  • Yong Dai — 6 papers, h 3
  • Yong Dai — 5 papers, h 4
  • Yong Dai — 3 papers, h 6
  • Yong Dai — 3 papers, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.CVShow all

4 papers · 1 filter

cs.CV2026

OPUS: A Simple yet Effective Unified Framework for Open-Vocabulary Detection

Xiaoyan Wei, Zhimin Yao, Ruilin Yang +4

Recent unified open-vocabulary detection (OVD) supports heterogeneous prompts, including text queries, visual exemplars, and their combinations, but often rely on increasingly comp…

cs.CV2026

Thinking Once Is Enough: Intermediate-Layer Evidence Routing for High-Resolution VQA

Zhongkuan Mao, Xianjie Liu, Tianyu Meng +9

High-resolution visual question answering (HR-VQA) is often treated as a problem of insufficient evidence acquisition, where failing multimodal large language models must inspect i…

cs.CV2026

Current World Models Lack a Persistent State Core

Jinpeng Lu, Dexu Zhu, Haoyuan Shi +8

World models are increasingly regarded as a decisive step toward artificial general intelligence, yet modeling the physical world demands more than rendering convincing frames on d…

cs.CV2026

EPIC-Bench: A Perception-Centric Benchmark for Fine-Grained Embodied Visual Grounding in Vision-Language Models

Haozhe Shan, Xiancong Ren, Han Dong +9

While large vision-language models (VLMs) are increasingly adopted as the perceptual backbone for embodied agents, existing benchmarks often rely on question-answering or multiple-…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.