activity
20162024
most citedVideo-P2P: Video Editing with Cross-attention Control

14 citations · 42 across the 6 of their papers we have counts for

collaborators

6 papers

cs.NI2024

DHNet: A Distributed Network Architecture for Smart Home

Chaoqi Zhou, Jingpu Duan, YuPeng Xiao +4

With the increasing popularity of smart homes, more and more devices need to connect to home networks. Traditional home networks mainly rely on centralized networking, where an exc…

cs.CV202412 cited

Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Yanwei Li, Yuechen Zhang, Chengyao Wang +5

In this work, we introduce Mini-Gemini, a simple and effective framework enhancing multi-modality Vision Language Models (VLMs). Despite the advancements in VLMs facilitating basic…

cs.CV20239 cited

Direct Inversion: Boosting Diffusion-based Editing with 3 Lines of Code

Xuan Ju, Ailing Zeng, Yuxuan Bian +2

Text-guided diffusion models have revolutionized image generation and editing, offering exceptional realism and diversity. Specifically, in the context of diffusion-based editing,…

cs.CV2023

Self-supervised Learning by View Synthesis

Shaoteng Liu, Xiangyu Zhang, Tao Hu +1

We present view-synthesis autoencoders (VSA) in this paper, which is a self-supervised learning framework designed for vision transformers. Different from traditional 2D pretrainin…

cs.CV202314 cited

Video-P2P: Video Editing with Cross-attention Control

Shaoteng Liu, Yuechen Zhang, Wenbo Li +2

This paper presents Video-P2P, a novel framework for real-world video editing with cross-attention control. While attention control has proven effective for image editing with pre-…

cs.NI20167 cited

Final Service Provider DevOps concept and evaluation

Guido Marchetto, Riccardo Sisto, Wolfgang John +19

This report presents the results of the UNIFY Service Provider DevOps activities. First, we present the final definition and assessment of the concept. SP-DevOps is realized by a c…