collaborators

9 papers

cs.CV2026

FACT: A Forensic Agent with Compiled Tool-Use Trajectories for AI-Generated Image Detection

Jiaoyang Chen, Bin Hu, Jingyu Hu +3

AI-generated image detection is increasingly open-world: new image generators produce highly realistic images that make visual artifacts harder to identify. Existing detectors usua…

cs.CV2026

FuncRoom-Agent: Sequential Feed-Forward 3D Functional Indoor Scene Generation

Hao Feng, Zhi Zuo, MingJian Liang +6

We introduce Function-Room Generation, a new indoor 3D scene generation setting that creates rooms supporting explicit functional goals rather than merely visually plausible layout…

cs.CV2026

SkelGen4D: Weakly-Supervised Skeleton-Based 4D Generation for Text-Driven Mesh Animation

Hao Feng, Zhi Zuo, Jia-Hui Pan +6

We study 4D generation to synthesize temporally coherent sequences of 3D geometry for animation and content creation. In contrast to existing SDS-based optimization methods and vid…

cs.CV2026

CamFlow+: Hybrid Motion Bases for 2D Camera Motion Estimation with Stabilization Applications

Haipeng Li, Zhen Liu, Zhanglei Yang +6

Estimating 2D camera motion is fundamental to computer vision and computational photography. Existing homography-based methods work well for planar scenes or pure rotation, but str…

cs.CV2026

SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images

Zishan Liu, Ruoxi Zang, Yanglin Zhang +5

Recent advancements in Large Vision-Language Models (VLMs) have demonstrated exceptional semantic understanding, yet these models consistently struggle with spatial reasoning, ofte…

cs.CV2026

PhyMix: Towards Physically Consistent Single-Image 3D Indoor Scene Generation with Implicit--Explicit Optimization

Dongli Wu, Jingyu Hu, Ka-Hei Hui +4

Existing single-image 3D indoor scene generators often produce results that look visually plausible but fail to obey real-world physics, limiting their reliability in robotics, emb…