9 papers · 1 filter
UniTraffic-Agent: Unified Traffic Video Reasoning for AI City Challenge 2026 Track 3 with Two Out-of-Domain Evaluations
Peng Li, Qianqian Xu, Shilong Bao +2
Traffic video understanding has become an important problem in intelligent transportation, as road videos provide direct evidence for accidents, violations, and interactions betwee…
Enhancing Object Coherence in Layout-to-Image Synthesis
Yibin Wang, Changhai Zhou, Honghui Xu
Layout-to-image synthesis is an emerging technique in conditional image generation. It aims to generate complex scenes, where users require fine control over the layout of the obje…
DreamText: High Fidelity Scene Text Synthesis
Yibin Wang, Weizhong Zhang, Honghui Xu +1
Scene text synthesis involves rendering specified texts onto arbitrary images. Current methods typically formulate this task in an end-to-end manner but lack effective character-le…
MagicFace: Training-free Universal-Style Human Image Customized Synthesis
Yibin Wang, Weizhong Zhang, Cheng Jin
Current human image customization methods leverage Stable Diffusion (SD) for its rich semantic prior. However, since SD is not specifically designed for human-oriented generation,…
PrimeComposer: Faster Progressively Combined Diffusion for Image Composition with Attention Steering
Yibin Wang, Weizhong Zhang, Jianwei Zheng +1
Image composition involves seamlessly integrating given objects into a specific visual context. Current training-free methods rely on composing attention weights from several sampl…
Out-of-Distribution Detection using Neural Activation Prior
Weilin Wan, Weizhong Zhang, Quan Zhou +2
Out-of-distribution detection (OOD) is a crucial technique for deploying machine learning models in the real world to handle the unseen scenarios. In this paper, we first propose a…