3 papers
cs.CV2025
LayerComposer: Multi-Human Personalized Generation via Layered Canvas
Guocheng Gordon Qian, Ruihang Zhang, Tsai-Shien Chen +10
Despite their impressive visual fidelity, existing personalized image generators lack interactive control over spatial composition and scale poorly to multiple humans. To address t…
cs.CV2025
Preventing Shortcuts in Adapter Training via Providing the Shortcuts
Anujraaj Argo Goyal, Guocheng Gordon Qian, Huseyin Coskun +8
Adapter-based training has emerged as a key mechanism for extending the capabilities of powerful foundation image generators, enabling personalized and stylized text-to-image synth…
cs.CV2025
One-Shot Medical Video Object Segmentation via Temporal Contrastive Memory Networks
Yaxiong Chen, Junjian Hu, Chunlei Li +6
Video object segmentation is crucial for the efficient analysis of complex medical video data, yet it faces significant challenges in data availability and annotation. We introduce…