◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Jia Li

9 papers hereh-index 4233 citations11 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author3

Across the 7 of 9 papers where every author was matched, so the position is known.

fields
  • cs.CV4
  • cs.CL2
  • cs.AI1
  • cs.HC1
  • cs.SD1
same name
  • Jia Li — 31 papers, h 16
  • Jia Li — 25 papers, h 28
  • Jia Li — 21 papers, h 5
  • Jia Li — 21 papers, h 8
  • Jia Li — 17 papers, h 3
  • Jia Li — 15 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedKimi K2.5: Visual Agentic Intelligence

2 citations · 2 across the 9 of their papers we have counts for

collaborators
Showing cs.CVShow all

4 papers · 1 filter

cs.CV2026

PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models

Zichao Lin, Yifeng Xie, Bowen Qu +30

We introduce PerceptionBench, a benchmark specifically designed to evaluate the atomic visual perception capabilities of Multimodal Large Language Models (MLLMs). Existing benchmar…

cs.CV2026

ARGaze: Autoregressive Transformers for Online Egocentric Gaze Estimation

Jia Li, Wenjie Zhao, Shijian Deng +6

Online egocentric gaze estimation predicts where a camera wearer is looking from first-person video using only past and current frames, a task essential for augmented reality and a…

cs.CV2025

Toward Gaze Target Detection of Young Autistic Children

Shijian Deng, Erin E. Kosloski, Siva Sai Nagender Vasireddy +6

The automatic detection of gaze targets in autistic children through artificial intelligence can be impactful, especially for those who lack access to a sufficient number of profes…

cs.CV2025

From Waveforms to Pixels: A Survey on Audio-Visual Segmentation

Jia Li, Yapeng Tian

Audio-Visual Segmentation (AVS) aims to identify and segment sound-producing objects in videos by leveraging both visual and audio modalities. It has emerged as a significant resea…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.