activity
20232026
most citedMulti-Prompts Learning with Cross-Modal Alignment for Attribute-based Person Re-Identification

2 citations · 3 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV2026

Tele-Omni: a Unified Multimodal Framework for Video Generation and Editing

Jialun Liu, Tian Li, Xiao Cao +20

Recent advances in diffusion-based video generation have substantially improved visual fidelity and temporal coherence. However, most existing approaches remain task-specific and r…

cs.CV2025

3DOT: Texture Transfer for 3DGS Objects from a Single Reference Image

Xiao Cao, Beibei Lin, Bo Wang +2

3D texture swapping allows for the customization of 3D object textures, enabling efficient and versatile visual transformations in 3D editing. While no dedicated method exists, ada…

cs.CV2024★ 1 cited

SSNeRF: Sparse View Semi-supervised Neural Radiance Fields with Augmentation

Xiao Cao, Beibei Lin, Bo Wang +2

Sparse view NeRF is challenging because limited input images lead to an under constrained optimization problem for volume rendering. Existing methods address this issue by relying…

cs.CL2024

Towards Better Text-to-Image Generation Alignment via Attention Modulation

Yihang Wu, Xiao Cao, Kaixin Li +4

In text-to-image generation tasks, the advancements of diffusion models have facilitated the fidelity of generated results. However, these models encounter challenges when processi…

cs.CV2023★ 2 cited

Multi-Prompts Learning with Cross-Modal Alignment for Attribute-based Person Re-Identification

Yajing Zhai, Yawen Zeng, Zhiyong Huang +3

The fine-grained attribute descriptions can significantly supplement the valuable semantic information for person image, which is vital to the success of person re-identification (…