2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.LG2026
RST: certifying and constructing prescribed information in variational autoencoders
Zegu Zhang, Jianhua Peng, Jian Zhang
Posterior-collapse diagnostics can show that a VAE uses its latent representation, but they do not determine whether a specified information view crosses a representation/readout i…
cs.CV2024
EMOdiffhead: Continuously Emotional Control in Talking Head Generation via Diffusion
Jian Zhang, Weijian Mai, Zhijun Zhang
The task of audio-driven portrait animation involves generating a talking head video using an identity image and an audio track of speech. While many existing approaches focus on l…
cs.AI2024★ 2 cited
Brain-Conditional Multimodal Synthesis: A Survey and Taxonomy
Weijian Mai, Jian Zhang, Pengfei Fang +1
In the era of Artificial Intelligence Generated Content (AIGC), conditional multimodal synthesis technologies (e.g., text-to-image, text-to-video, text-to-audio, etc) are gradually…