activity
20242026
most citedWISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

4 citations · 4 across the 2 of their papers we have counts for

collaborators

5 papers

cs.CV2026

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation

Juncheng Ma, Yuxuan Du, Yanan Sun +8

Diffusion Transformers (DiTs) have significantly advanced audio-driven portrait animation, but their high computational cost leads to substantial inference latency. Although traini…

cs.CV20264 cited

WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

Yuwei Niu, Munan Ning, Mengren Zheng +9

Text-to-Image (T2I) models are capable of generating high-quality artistic creations and visual content. However, existing research and evaluation standards predominantly focus on…

cs.NE2025

Event-based backpropagation on the neuromorphic platform SpiNNaker2

Gabriel Béna, Timo Wunderlich, Mahmoud Akl +3

Neuromorphic computing aims to replicate the brain's capabilities for energy efficient and parallel information processing, promising a solution to the increasing demand for faster…

cs.CV2025

SwapAnyone: Consistent and Realistic Video Synthesis for Swapping Any Person into Any Video

Chengshu Zhao, Yunyang Ge, Xinhua Cheng +6

Video body-swapping aims to replace the body in an existing video with a new body from arbitrary sources, which has garnered more attention in recent years. Existing methods treat…

cs.CV2024

DreamDance: Animating Human Images by Enriching 3D Geometry Cues from 2D Poses

Yatian Pang, Bin Zhu, Bin Lin +5

In this work, we present DreamDance, a novel method for animating human images using only skeleton pose sequences as conditional inputs. Existing approaches struggle with generatin…