Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Delayed Bidirectional Alignment via Disentangled Audio Semantics for Audio-Visual Segmentation
Jingqi Tian, Yiheng Du, Haoji Zhang +6
Audio-Visual Segmentation (AVS) aims to localize sound-producing objects at the pixel level by integrating auditory and visual cues. However, existing methods often struggle with m…
cs.CV2025
AlignedGen: Aligning Style Across Generated Images
Jiexuan Zhang, Yiheng Du, Qian Wang +3
Despite their generative power, diffusion models struggle to maintain style consistency across images conditioned on the same style prompt, hindering their practical deployment in…