1 paper
Chuyan Chen, Haoxing Chen, Kun Chen +27
We introduce LLaDA-Image, a unified framework that pairs a 6B Diffusion Transformer (DiT) trained from scratch with a frozen vision-language understanding module built on the LLaDA…