papers

Publications (35)

cs.HC2024

LAVE: LLM-Powered Agent Assistance and Language Augmentation for Video Editing

Bryan Wang, Yuliang Li, Zhaoyang Lv +3

Video creation has become increasingly popular, yet the expertise and effort required for editing often pose barriers to beginners. In this paper, we explore the integration of lar…

cs.HC2025

Script2Screen: Supporting Dialogue Scriptwriting with Interactive Audiovisual Generation

Zhecheng Wang, Jiaju Ma, Eitan Grinspun +2

Scriptwriting has traditionally been text-centric, a modality that only partially conveys the produced audiovisual experience. A formative study with professional writers informed…

cs.HC2024

SimTube: Generating Simulated Video Comments through Multimodal AI and User Personas

Yu-Kai Hung, Yun-Chien Huang, Ting-Yu Su +4

Audience feedback is crucial for refining video content, yet it typically comes after publication, limiting creators' ability to make timely adjustments. To bridge this gap, we int…

cs.HC2026

Vidmento: Creating Video Stories Through Context-Aware Expansion With Generative Video

Catherine Yeh, Anh Truong, Mira Dontcheva +1

Video storytelling is often constrained by available material, limiting creative expression and leaving undesired narrative gaps. Generative video offers a new way to address these…

cs.CV2023

Hierarchical Conditional Semi-Paired Image-to-Image Translation For Multi-Task Image Defect Correction On Shopping Websites

Moyan Li, Jinmiao Fu, Shaoyuan Xu +3

On shopping websites, product images of low quality negatively affect customer experience. Although there are plenty of work in detecting images with different defects, few efforts…

cs.LG2023

KD-FixMatch: Knowledge Distillation Siamese Neural Networks

Chien-Chih Wang, Shaoyuan Xu, Jinmiao Fu +2

Semi-supervised learning (SSL) has become a crucial approach in deep learning as a way to address the challenge of limited labeled data. The success of deep neural networks heavily…