3 papers
cs.CV2024
UVCG: Leveraging Temporal Consistency for Universal Video Protection
KaiZhou Li, Jindong Gu, Xinchun Yu +3
The security risks of AI-driven video editing have garnered significant attention. Although recent studies indicate that adding perturbations to images can protect them from malici…
cs.CV2024
Localizing Events in Videos with Multimodal Queries
Gengyuan Zhang, Mang Ling Ada Fok, Jialu Ma +5
Localizing events in videos based on semantic queries is a pivotal task in video understanding, with the growing significance of user-oriented applications like video search. Yet,…
cs.CY2024
OpenCarbonEval: A Unified Carbon Emission Estimation Framework in Large-Scale AI Models
Zhaojian Yu, Yinghao Wu, Zhuotao Deng +2
In recent years, large-scale auto-regressive models have made significant progress in various tasks, such as text or video generation. However, the environmental impact of these mo…