◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Tianyi Zhou

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV3
  • cs.CL1
ORCID 0000-0002-1488-8122
same name
  • Tianyi Zhou — 45 papers
  • Tianyi Zhou — 22 papers, h 33
  • Tianyi Zhou — 2 papers
  • Tianyi Zhou — 2 papers
  • Tianyi Zhou — 2 papers
  • Tianyi Zhou — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedBLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

1 citations · 1 across the 4 of their papers we have counts for

collaborators

4 papers

cs.CV2025

BLIP3o-NEXT: Next Frontier of Native Image Generation

Jiuhai Chen, Le Xue, Zhiyang Xu +12

We present BLIP3o-NEXT, a fully open-source foundation model in the BLIP3 series that advances the next frontier of native image generation. BLIP3o-NEXT unifies text-to-image gener…

cs.CL2025

Submodular Context Partitioning and Compression for In-Context Learning

Shaoyi Zheng, Canyu Zhang, Tianyi Zhou +1

In-context learning (ICL) enables efficient few-shot learning in large language models (LLMs) without training, but suffers from the quadratic input complexity of transformers, lim…

cs.CV2025★ 1 cited

BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Jiuhai Chen, Zhiyang Xu, Xichen Pan +10

Unifying image understanding and generation has gained growing attention in recent research on multimodal models. Although design choices for image understanding have been extensiv…

cs.CV2024

Florence-VL: Enhancing Vision-Language Models with Generative Vision Encoder and Depth-Breadth Fusion

Jiuhai Chen, Jianwei Yang, Haiping Wu +4

We present Florence-VL, a new family of multimodal large language models (MLLMs) with enriched visual representations produced by Florence-2, a generative vision foundation model.…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.