Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
InsightTok: Improving Text and Face Fidelity in Discrete Tokenization for Autoregressive Image Generation
Yang Yue, Fangyun Wei, Tianyu He +10
Text and faces are among the most perceptually salient and practically important patterns in visual generation, yet they remain challenging for autoregressive generators built on d…
cs.CV2024
DiLightNet: Fine-grained Lighting Control for Diffusion-based Image Generation
Chong Zeng, Yue Dong, Pieter Peers +3
This paper presents a novel method for exerting fine-grained lighting control during text-driven diffusion-based image generation. While existing diffusion models already have the…