◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Manling Li

10 papers hereh-index 16958 citations45 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author10

Across the 10 of 10 papers where every author was matched, so the position is known.

fields
  • cs.CV4
  • cs.AI3
  • cs.CL1
  • cs.SE1
  • cs.SI1
same name
  • Manling Li — 12 papers, h 9
  • Manling Li — 8 papers, h 3
  • Manling Li — 5 papers
  • Manling Li — 5 papers, h 4
  • Manling Li — 4 papers, h 3
  • Manling Li — 3 papers, h 13

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20212025
most citedLearning to Decompose Visual Features with Latent Textual Prompts

8 citations · 13 across the 6 of their papers we have counts for

collaborators
Showing cs.CVShow all

4 papers · 1 filter

cs.CV2023

ViStruct: Visual Structural Knowledge Extraction via Curriculum Guided Code-Vision Representation

Yangyi Chen, Xingyao Wang, Manling Li +2

State-of-the-art vision-language models (VLMs) still have limited performance in structural knowledge extraction, such as relations between objects. In this work, we present ViStru…

cs.CV2022★ 2 cited

Video Event Extraction via Tracking Visual States of Arguments

Guang Yang, Manling Li, Jiajie Zhang +3

Video event extraction aims to detect salient events from a video and identify the arguments for each event as well as their semantic roles. Existing methods focus on capturing the…

cs.CV2022★ 8 cited

Learning to Decompose Visual Features with Latent Textual Prompts

Feng Wang, Manling Li, Xudong Lin +3

Recent advances in pre-training vision-language models like CLIP have shown great potential in learning transferable visual representations. Nonetheless, for downstream inference,…

cs.CV2021★ 3 cited

Joint Multimedia Event Extraction from Video and Article

Brian Chen, Xudong Lin, Christopher Thomas +5

Visual and textual modalities contribute complementary information about events described in multimedia documents. Videos contain rich dynamics and detailed unfoldings of events, w…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.