9 citations · 14 across the 3 of their papers we have counts for
6 papers
LDGen: Enhancing Text-to-Image Synthesis via Large Language Model-Driven Language Representation
Pengzhi Li, Pengfei Yu, Zide Liu +5
In this paper, we introduce LDGen, a novel method for integrating large language models (LLMs) into existing text-to-image diffusion models while minimizing computational demands.…
ART3D: 3D Gaussian Splatting for Text-Guided Artistic Scenes Generation
Pengzhi Li, Chengshuai Tang, Qinxuan Huang +1
In this paper, we explore the existing challenges in 3D artistic scene generation by introducing ART3D, a novel framework that combines diffusion models and 3D Gaussian splatting t…
Generating Daylight-driven Architectural Design via Diffusion Models
Pengzhi Li, Baijuan Li
In recent years, the rapid development of large-scale models has made new possibilities for interdisciplinary fields such as architecture. In this paper, we present a novel dayligh…
The Devil is in the Edges: Monocular Depth Estimation with Edge-aware Consistency Fusion
Pengzhi Li, Yikang Ding, Haohan Wang +2
This paper presents a novel monocular depth estimation method, named ECFNet, for estimating high-quality monocular depth with clear edges and valid overall structure from a single…
Sketch-to-Architecture: Generative AI-aided Architectural Design
Pengzhi Li, Baijuan Li, Zhiheng Li
Recently, the development of large-scale models has paved the way for various interdisciplinary research, including architecture. By using generative AI, we present a novel workflo…
Tuning-Free Image Customization with Image and Text Guidance
Pengzhi Li, Qiang Nie, Ying Chen +7
Despite significant advancements in image customization with diffusion models, current methods still have several limitations: 1) unintended changes in non-target areas when regene…