3 papers
cs.CV2025
LDGen: Enhancing Text-to-Image Synthesis via Large Language Model-Driven Language Representation
Pengzhi Li, Pengfei Yu, Zide Liu +5
In this paper, we introduce LDGen, a novel method for integrating large language models (LLMs) into existing text-to-image diffusion models while minimizing computational demands.…
cs.CV2024
ART3D: 3D Gaussian Splatting for Text-Guided Artistic Scenes Generation
Pengzhi Li, Chengshuai Tang, Qinxuan Huang +1
In this paper, we explore the existing challenges in 3D artistic scene generation by introducing ART3D, a novel framework that combines diffusion models and 3D Gaussian splatting t…
cs.CV2024
Generating Daylight-driven Architectural Design via Diffusion Models
Pengzhi Li, Baijuan Li
In recent years, the rapid development of large-scale models has made new possibilities for interdisciplinary fields such as architecture. In this paper, we present a novel dayligh…