2 papers
cs.CV2025
Text-to-Edit: Controllable End-to-End Video Ad Creation via Multimodal LLMs
Dabing Cheng, Haosen Zhan, Xingchen Zhao +6
The exponential growth of short-video content has ignited a surge in the necessity for efficient, automated solutions to video editing, with challenges arising from the need to und…
cs.CV2023
Unsupervised Domain Adaptation for Semantic Segmentation with Pseudo Label Self-Refinement
Xingchen Zhao, Niluthpol Chowdhury Mithun, Abhinav Rajvanshi +2
Deep learning-based solutions for semantic segmentation suffer from significant performance degradation when tested on data with different characteristics than what was used during…