◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Chengzu Li

University of Cambridge

24 papers hereh-index 111.2k citations29 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author6
  • middle author18

Across the 24 of 24 papers where every author was matched, so the position is known.

fields
  • cs.CL14
  • cs.CV7
  • cs.LG3
affiliations
  • University of Cambridge
Homepage

identity via Semantic Scholar / OpenAlex

activity
20222026
most citedBinding Language Models in Symbolic Languages

38 citations · 41 across the 23 of their papers we have counts for

collaborators
Showing cs.LGShow all

3 papers · 1 filter

cs.LG2026

Thinking in Frames: How Visual Context and Test-Time Scaling Empower Video Reasoning

Chengzu Li, Zanyi Wang, Jiaang Li +9

Vision-Language Models have excelled at textual reasoning, but they often struggle with fine-grained spatial understanding and continuous action planning, failing to simulate the d…

cs.LG2025

Visual Planning: Let's Think Only with Images

Yi Xu, Chengzu Li, Han Zhou +4

Recent advancements in Large Language Models (LLMs) and their multimodal extensions (MLLMs) have substantially enhanced machine reasoning across diverse tasks. However, these model…

cs.LG2025

Scaling and Beyond: Advancing Spatial Reasoning in MLLMs Requires New Recipes

Huanyu Zhang, Chengzu Li, Wenshan Wu +8

Multimodal Large Language Models (MLLMs) have demonstrated impressive performance in general vision-language tasks. However, recent studies have exposed critical limitations in the…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.