42 citations · 64 across the 23 of their papers we have counts for
33 papers
LILAC: Layer-Wise Independent LoRAs and Cascaded Conditioning for Multi-Concept Customization of Diffusion Models
Marian Lupascu, Sebastian Ripa, Mihai Trascau +2
Personalizing text-to-image diffusion models to render several specific subjects in a coherent image remains challenging: the model must preserve each subject's identity while keep…
LaDe: Unified Multi-Layered Graphic Media Generation and Decomposition
Vlad-Constantin Lungu-Stan, Ionut Mironica, Mariana-Iuliana Georgescu
Media design layer generation enables the creation of fully editable, layered design documents such as posters, flyers, and logos using only natural language prompts. Existing meth…
X-Aligner: Composed Visual Retrieval without the Bells and Whistles
Yuqian Zheng, Mariana-Iuliana Georgescu
Composed Video Retrieval (CoVR) facilitates video retrieval by combining visual and textual queries. However, existing CoVR frameworks typically fuse multimodal inputs in a single…
Road Obstacle Video Segmentation
Shyam Nandan Rai, Shyamgopal Karthik, Mariana-Iuliana Georgescu +3
With the growing deployment of autonomous driving agents, the detection and segmentation of road obstacles have become critical to ensure safe autonomous navigation. However, exist…
Q-Former Autoencoder: A Modern Framework for Medical Anomaly Detection
Francesco Dalmonte, Emirhan Bayar, Emre Akbas +1
Anomaly detection in medical images is an important yet challenging task due to the diversity of possible anomalies and the practical impossibility of collecting comprehensively an…
Subspace-Boosted Model Merging
Ronald Skorobogat, Karsten Roth, Mariana-Iuliana Georgescu
Model merging enables the combination of multiple specialized expert models into a single model capable of performing multiple tasks. However, the benefits of merging an increasing…