3d-free synthesis 1camera control 1self-supervised learning 1text-driven viewpoint 1video re-shooting 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CV2026
TARS: Timestep-Aware Data Scaling for 3D-Free Video Re-Shooting
Jiwen Liu, Shujuan Li, Xiaohan Li +5
The paper introduces TARS, a 3D‑free video re‑shooting framework that uses text‑driven semantic viewpoint specifications and self‑supervised training to control camera motion and p…
cs.CV2026
Vera: Identity-Faithful Human Subject-to-Video Generation
Yulong Xu, Xinyue Liu, Shujuan Li +6
Subject-to-video (S2V) generation has made substantial progress in preserving reference subjects across diverse categories, yet generic subject consistency remains insufficient for…
cs.CV2026
MVPBench: A Multi-Video Perception Evaluation Benchmark for Multi-Modal Video Understanding
Purui Bai, Tao Wu, Jiayang Sun +3
The rapid progress of Large Language Models (LLMs) has spurred growing interest in Multi-modal LLMs (MLLMs) and motivated the development of benchmarks to evaluate their perceptual…