◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Kai Han

5 papers hereh-index 12 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author3

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.CV3
  • cs.CL2
same name
  • Kai Han — 21 papers, h 36
  • Kai Han — 17 papers, h 15
  • Kai Han — 12 papers, h 3
  • Kai Han — 10 papers, h 5
  • Kai Han — 7 papers, h 2
  • Kai Han — 6 papers, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.CVShow all

3 papers · 1 filter

cs.CV2026

Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models

Xinpeng Dong, Min Zhang, Kairong Han +3

In recent years, multimodal large language models (MLLMs) have achieved remarkable progress, primarily attributed to effective paradigms for integrating visual and textual informat…

cs.CV2026

iVGR: Internalizing Visually Grounded Reasoning for MLLMs with Reinforcement Learning

Chang-Bin Zhang, Yujie Zhong, Qiang Zhang +1

While visually grounded Chain-of-Thought (CoT) has emerged as a promising paradigm to enhance fine-grained perception in multimodal large language models (MLLMs), its efficacy duri…

cs.CV2026

CodeBind: Decoupled Representation Learning for Multimodal Alignment with Unified Compositional Codebook

Zeyu Chen, Jie Li, Kai Han

Multimodal representation alignment is pivotal for large language models and robotics. Traditional methods are often hindered by cross-modal information discrepancies and data scar…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.