From the 1 of 4 linked papers with an AI index.
4 papers
VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System
Haodong Li, Tianfei Ren, Xiaoxiao Ma +25
The paper presents VideoCoCo, a system that generates physically consistent videos by having a coding agent produce executable Blender code that defines the scene and its dynamics,…
Unpaired Joint Distribution Modeling via Multi-Scale Image Representations
Yihang Zou, Hui Zhang, Zuowei Shen +1
This paper studies the problem of learning a joint distribution from marginal observations, which is inherently ill-posed due to the ambiguity of feasible couplings. We propose LUD…
CoCo: Code as CoT for Text-to-Image Preview and Rare Concept Generation
Haodong Li, Chunmei Qing, Huanyu Zhang +11
Recent advancements in Unified Multimodal Models (UMMs) have significantly advanced text-to-image (T2I) generation, particularly through the integration of Chain-of-Thought (CoT) r…
Enhancing Low-resolution Image Representation Through Normalizing Flows
Chenglong Bao, Tongyao Pang, Zuowei Shen +2
Low-resolution image representation is a special form of sparse representation that retains only low-frequency information while discarding high-frequency components. This property…