◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Kent Yu

4 papers hereh-index 3256 citations4 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV4

identity via Semantic Scholar / OpenAlex

activity
20232026
most citedT-Rex: Counting by Visual Prompting

5 citations · 5 across the 2 of their papers we have counts for

collaborators
Showing cs.CVShow all

3 papers · 1 filter

cs.CV2026

Reflective VLA: In-Context Action Consequences Make VLAs Generalize

Qing Lian, Kent Yu, Lei Zhang

Most vision-language-action (VLA) models are reactive: they predict the next action from the current instruction and observation, implicitly assuming that the current observation f…

cs.CV2024

DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding

Tianhe Ren, Yihao Chen, Qing Jiang +17

In this paper, we introduce DINO-X, which is a unified object-centric vision model developed by IDEA Research with the best open-world object detection performance to date. DINO-X…

cs.CV2023★ 5 cited

T-Rex: Counting by Visual Prompting

Qing Jiang, Feng Li, Tianhe Ren +4

We introduce T-Rex, an interactive object counting model designed to first detect and then count any objects. We formulate object counting as an open-set object detection task with…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.