activity
20242026
collaborators
Showing cs.CVShow all

8 papers · 1 filter

cs.CV2026

CameraAnything: Refilming Videos with Arbitrary Camera Control

Yixuan Li, Yanhong Zeng, Ka Leong Cheng +11

We introduce CameraAnything, the first unified framework for camera controlled video editing that enables joint control of both intrinsic and extrinsic camera parameters. Existing…

cs.CV2026

InterActHuman: Multi-Concept Human Animation with Layout-Aligned Audio Conditions

Zhenzhi Wang, Jiaqi Yang, Jianwen Jiang +7

End-to-end human animation with rich multi-modal conditions, e.g., text, image and audio has achieved remarkable advancements in recent years. However, most existing methods could…

cs.CV2025

Multi-identity Human Image Animation with Structural Video Diffusion

Zhenzhi Wang, Yixuan Li, Yanhong Zeng +4

Generating human videos from a single image while ensuring high visual quality and precise control is a challenging task, especially in complex scenarios involving multiple individ…

cs.CV2025

AnySplat: Feed-forward 3D Gaussian Splatting from Unconstrained Views

Lihan Jiang, Yucheng Mao, Linning Xu +9

We introduce AnySplat, a feed forward network for novel view synthesis from uncalibrated image collections. In contrast to traditional neural rendering pipelines that demand known…

cs.CV2025

Long Context Tuning for Video Generation

Yuwei Guo, Ceyuan Yang, Ziyan Yang +5

Recent advances in video generation can produce realistic, minute-long single-shot videos with scalable diffusion transformers. However, real-world narrative videos require multi-s…

cs.CV2024

Proc-GS: Procedural Building Generation for City Assembly with 3D Gaussians

Yixuan Li, Xingjian Ran, Linning Xu +6

Buildings are primary components of cities, often featuring repeated elements such as windows and doors. Traditional 3D building asset creation is labor-intensive and requires spec…