◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Kevin Qinghong Lin

22 papers hereh-index 121.4k citations24 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author17

Across the 21 of 22 papers where every author was matched, so the position is known.

fields
  • cs.CV20
  • cs.AI1
  • cs.RO1

identity via Semantic Scholar / OpenAlex

activity
20232026
most citedShow-o: One Single Transformer to Unify Multimodal Understanding and Generation

3 citations · 14 across the 21 of their papers we have counts for

collaborators
Showing 2026Show all

4 papers · 1 filter

cs.RO2026

Show-Harness: Just a VLM Agent Can Play Robots

Yanzhe Chen, Zechen Bai, Zhijun Cao +7

Foundation vision-language models (VLMs) exhibit broad intelligence about the world, yet translating this intelligence into robot control remains challenging. We present Show-Harne…

cs.CV2026

SurgNarrator: A Generative Retrieval Framework for Surgical Video Understanding

Yuqing Feng, Jiawei Ma, Kevin Qinghong Lin +6

Surgical procedures unfold as structured and recurring clinical events, whose real-time understanding via intraoperative surgical videos is critical for intraoperative decision-mak…

cs.CV2026

Demo2Tutorial: From Human Experience to Multimodal Software Tutorials

Zechen Bai, Zhiheng Chen, Yiqi Lin +5

Human experience in digital environments offers a vast, underexplored resource of authentic, untrimmed interactions that contain rich procedural knowledge. We introduce Demo2Tutori…

cs.CV2026

FocusUI: Efficient UI Grounding via Position-Preserving Visual Token Selection

Mingyu Ouyang, Kevin Qinghong Lin, Mike Zheng Shou +1

Vision-Language Models (VLMs) have shown remarkable performance in User Interface (UI) grounding tasks, driven by their ability to process increasingly high-resolution screenshots.…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.