4 papers · 1 filter
GHOST 2.0: generative high-fidelity one shot transfer of heads
Alexander Groshev, Anastasiia Iashchenko, Pavel Paramonov +2
While the task of face swapping has recently gained attention in the research community, a related problem of head swapping remains largely unexplored. In addition to skin color tr…
Kandinsky 3: Text-to-Image Synthesis for Multifunctional Generative Framework
Vladimir Arkhipkin, Viacheslav Vasilev, Andrei Filatov +9
Text-to-image (T2I) diffusion models are popular for introducing image manipulation methods, such as editing, image fusion, inpainting, etc. At the same time, image-to-video (I2V)…
Kandinsky 3.0 Technical Report
Vladimir Arkhipkin, Andrei Filatov, Viacheslav Vasilev +6
We present Kandinsky 3.0, a large-scale text-to-image generation model based on latent diffusion, continuing the series of text-to-image Kandinsky models and reflecting our progres…
OmniFusion Technical Report
Elizaveta Goncharova, Anton Razzhigaev, Matvey Mikhalchuk +6
Last year, multimodal architectures served up a revolution in AI-based approaches and solutions, extending the capabilities of large language models (LLM). We propose an \textit{Om…