◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Zhuoyi Yang

12 papers hereh-index 123.9k citations16 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author8
  • last author1

Across the 11 of 12 papers where every author was matched, so the position is known.

fields
  • cs.CV8
  • cs.CL2
  • stat.ME1
  • stat.ML1
same name
  • Zhuoyi Yang — 7 papers, h 6
  • Zhuoyi Yang — 6 papers, h 3
  • Zhuoyi Yang — 3 papers, h 0

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20182025
most citedCogView: Mastering Text-to-Image Generation via Transformers

383 citations · 770 across the 10 of their papers we have counts for

collaborators
Showing 2024Show all

4 papers · 1 filter

cs.CV2024

VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation

Jiazheng Xu, Yu Huang, Jiale Cheng +19

Visual generative models have achieved remarkable progress in synthesizing photorealistic images and videos, yet aligning their outputs with human preferences across critical dimen…

cs.CV2024★ 7 cited

CogVLM2: Visual Language Models for Image and Video Understanding

Wenyi Hong, Weihan Wang, Ming Ding +22

Beginning with VisualGLM and CogVLM, we are continuously exploring VLMs in pursuit of enhanced vision-language fusion, efficient higher-resolution architecture, and broader modalit…

cs.CV2024

Inf-DiT: Upsampling Any-Resolution Image with Memory-Efficient Diffusion Transformer

Zhuoyi Yang, Heyang Jiang, Wenyi Hong +5

Diffusion models have shown remarkable performance in image generation in recent years. However, due to a quadratic increase in memory during generating ultra-high-resolution image…

cs.CV2024

CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Wendi Zheng, Jiayan Teng, Zhuoyi Yang +6

Recent advancements in text-to-image generative systems have been largely driven by diffusion models. However, single-stage text-to-image diffusion models still face challenges, in…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.