activity
20242026
collaborators

5 papers

cs.CV2026

Learning to Generate Rigid Body Interactions with Video Diffusion Models

David Romero, Ariana Bermudez, Viacheslav Iablochnikov +3

Recent video generation models have achieved remarkable progress and are now deployed in film, social media production, and advertising. Beyond their creative potential, such model…

cs.CV2025

MFM-point: Multi-scale Flow Matching for Point Cloud Generation

Petr Molodyk, Jaemoo Choi, David W. Romero +2

In recent years, point cloud generation has gained significant attention in 3D generative modeling. Among existing approaches, point-based methods directly generate point clouds wi…

cs.CV2025

HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation

Hermann Kumbong, Xian Liu, Tsung-Yi Lin +6

Visual Auto-Regressive modeling (VAR) has shown promise in bridging the speed and quality gap between autoregressive image models and diffusion models. VAR reformulates autoregress…

cs.AI2025

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

NVIDIA, :, Alisson Azzolini +51

Physical AI systems need to perceive, understand, and perform complex actions in the physical world. In this paper, we present the Cosmos-Reason1 models that can understand the phy…

cs.GR2024

Meshtron: High-Fidelity, Artist-Like 3D Mesh Generation at Scale

Zekun Hao, David W. Romero, Tsung-Yi Lin +1

Meshes are fundamental representations of 3D surfaces. However, creating high-quality meshes is a labor-intensive task that requires significant time and expertise in 3D modeling.…