Publications (35)
LAVE: LLM-Powered Agent Assistance and Language Augmentation for Video Editing
Bryan Wang, Yuliang Li, Zhaoyang Lv +3
Video creation has become increasingly popular, yet the expertise and effort required for editing often pose barriers to beginners. In this paper, we explore the integration of lar…
Script2Screen: Supporting Dialogue Scriptwriting with Interactive Audiovisual Generation
Zhecheng Wang, Jiaju Ma, Eitan Grinspun +2
Scriptwriting has traditionally been text-centric, a modality that only partially conveys the produced audiovisual experience. A formative study with professional writers informed…
SimTube: Generating Simulated Video Comments through Multimodal AI and User Personas
Yu-Kai Hung, Yun-Chien Huang, Ting-Yu Su +4
Audience feedback is crucial for refining video content, yet it typically comes after publication, limiting creators' ability to make timely adjustments. To bridge this gap, we int…
Vidmento: Creating Video Stories Through Context-Aware Expansion With Generative Video
Catherine Yeh, Anh Truong, Mira Dontcheva +1
Video storytelling is often constrained by available material, limiting creative expression and leaving undesired narrative gaps. Generative video offers a new way to address these…
Hierarchical Conditional Semi-Paired Image-to-Image Translation For Multi-Task Image Defect Correction On Shopping Websites
Moyan Li, Jinmiao Fu, Shaoyuan Xu +3
On shopping websites, product images of low quality negatively affect customer experience. Although there are plenty of work in detecting images with different defects, few efforts…
KD-FixMatch: Knowledge Distillation Siamese Neural Networks
Chien-Chih Wang, Shaoyuan Xu, Jinmiao Fu +2
Semi-supervised learning (SSL) has become a crucial approach in deep learning as a way to address the challenge of limited labeled data. The success of deep neural networks heavily…