activity
20242026
most citedUncovering the Text Embedding in Text-to-Image Diffusion Models

1 citations · 2 across the 5 of their papers we have counts for

collaborators
Showing cs.CVShow all

8 papers · 1 filter

cs.CV2026

UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception

Qin Guo, Hao Luo, Dongxu Yue +4

Recent advances in diffusion models have shown impressive performance in controllable image generation and dense prediction tasks. However, existing approaches typically treat diff…

cs.CV2025

Preacher: Paper-to-Video Agentic System

Jingwei Liu, Ling Yang, Hao Luo +3

The paper-to-video task converts a research paper into a structured video abstract, distilling key concepts, methods, and conclusions into an accessible, well-organized format. Whi…

cs.CV2025

SynBoost: A Synergistic Framework for Fast Sampling of Diffusion Models

Hu Yu, Hao Luo, Xueyang Fu +3

Diffusion probabilistic models (DPMs) have demonstrated remarkable success in visual generation. However, their iterative sampling mechanism results in slow inference speeds. While…

cs.CV2025

PlayerOne: Egocentric World Simulator

Yuanpeng Tu, Hao Luo, Xi Chen +3

We introduce PlayerOne, the first egocentric realistic world simulator, facilitating immersive and unrestricted exploration within vividly dynamic environments. Given an egocentric…

cs.CV2024

AnyLogo: Symbiotic Subject-Driven Diffusion System with Gemini Status

Jinghao Zhang, Wen Qian, Hao Luo +2

Diffusion models have made compelling progress on facilitating high-throughput daily production. Nevertheless, the appealing customized requirements are remain suffered from instan…

cs.CV20241 cited

Uncovering the Text Embedding in Text-to-Image Diffusion Models

Hu Yu, Hao Luo, Fan Wang +1

The correspondence between input text and the generated image exhibits opacity, wherein minor textual modifications can induce substantial deviations in the generated image. While,…