From the 1 of 19 linked papers with an AI index.
19 papers
Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Mesh
Junhao Chen, Mingjin Chen, Henghaofan Zhang +10
Pretrained video diffusion models can act as renderers when the desired scene state is already specified by an animated mesh, a camera trajectory, and a reference image. This 4D ge…
Bunraku: Turning a Single Illustration into an Editable Live2D Character
Junhao Chen, Jingjia Mao, Dayong Li +6
The paper introduces Bunraku, a system that automatically creates a complete Live2D character—including layered RGBA images, deformation meshes, and animation keyposes—from a singl…
GeoStereo: A Unified Stereo Geometry Estimation Framework for Disparity and Surface Normal
Qizhe Wei, Xianda Guo, Shaocong Xu +3
Stereo matching and surface normal estimation are fundamental tasks in 3D vision. However, existing feed-forward stereo methods still struggle to produce reliable predictions in ch…
Engine-Native Editable 3D World Reconstruction with Objects and Lighting
Junhao Chen, Xinghao Chen, Henghaofan Zhang +8
Editable 3D scene creation requires object instances and lights that can be inspected, moved, and imported into standard engines, yet existing single-image methods largely stop at…
Feedforward 3D Editing Learns from Semantic-Part Transformation
Jiawei Weng, Saining Zhang, Zhenxin Diao +4
3D editing is a fundamental capability for scalable 3D content creation. While image editing has rapidly evolved toward large-scale feedforward generative paradigms, 3D AI generati…
One Video, One World: Turning Monocular Video into Physical 4D Scenes
Junhao Chen, Boran Zhang, Mingjin Chen +7
We introduce \textbf{OVOW}, the first training-free system that reconstructs \emph{instance-level, simulation-ready} 4D mesh scenes from a single monocular video. Recent 4D reconst…