3 papers
cs.CV2026
Thermal-Only Crowd Counting with Deployment-Time Privacy Protection
Yifei Qian, Zhongliang Guo, Chun Tong Lei +4
While RGB-Thermal crowd counting has shown promise, the paradigm faces critical limitations: RGB data raises privacy concerns in public surveillance, and multi-modal misalignment d…
cs.CV2026
Uni-Animator: Towards Unified Visual Colorization
Xinyuan Chen, Yao Xu, Shaowen Wang +2
We propose Uni-Animator, a novel Diffusion Transformer (DiT)-based framework for unified image and video sketch colorization. Existing sketch colorization methods struggle to unify…
cs.CV2025
T2ICount: Enhancing Cross-modal Understanding for Zero-Shot Counting
Yifei Qian, Zhongliang Guo, Bowen Deng +5
Zero-shot object counting aims to count instances of arbitrary object categories specified by text descriptions. Existing methods typically rely on vision-language models like CLIP…