3 papers
cs.CV2026
JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation
Yinan Chen, Chuming Lin, Zhennan Chen +12
While instruction-based video editing has seen significant progress, joint audio-visual editing remains constrained by the absence of dedicated datasets and benchmarks. To bridge t…
cs.CL2024
Synth-Empathy: Towards High-Quality Synthetic Empathy Data
Hao Liang, Linzhuang Sun, Jingxuan Wei +5
In recent years, with the rapid advancements in large language models (LLMs), achieving excellent empathetic response capabilities has become a crucial prerequisite. Consequently,…
cs.CV2024
KeyVideoLLM: Towards Large-scale Video Keyframe Selection
Hao Liang, Jiapeng Li, Tianyi Bai +7
Recently, with the rise of web videos, managing and understanding large-scale video datasets has become increasingly important. Video Large Language Models (VideoLLMs) have emerged…