activity
20142023
most citedDebiasing Vision-Language Models via Biased Prompts

22 citations · 75 across the 17 of their papers we have counts for

collaborators

30 papers

cs.CV2024

CompGS: Unleashing 2D Compositionality for Compositional Text-to-3D via Dynamically Optimizing 3D Gaussians

Chongjian Ge, Chenfeng Xu, Yuanfeng Ji +6

Recent breakthroughs in text-guided image generation have significantly advanced the field of 3D generation. While generating a single high-quality 3D object is now feasible, gener…

cs.CV20244 cited

SF3D: Stable Fast 3D Mesh Reconstruction with UV-unwrapping and Illumination Disentanglement

Mark Boss, Zixuan Huang, Aaryaman Vasishta +1

We present SF3D, a novel method for rapid and high-quality textured object mesh reconstruction from a single image in just 0.5 seconds. Unlike most existing approaches, SF3D is exp…

cs.CV2024

SMooDi: Stylized Motion Diffusion Model

Lei Zhong, Yiming Xie, Varun Jampani +2

We introduce a novel Stylized Motion Diffusion model, dubbed SMooDi, to generate stylized motion driven by content texts and style motion sequences. Unlike existing methods that ei…

cs.CV2024

SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing

Ruihuang Li, Liyi Chen, Zhengqiang Zhang +3

Text-based 2D diffusion models have demonstrated impressive capabilities in image generation and editing. Meanwhile, the 2D diffusion models also exhibit substantial potentials for…

cs.CV2024

ICE-G: Image Conditional Editing of 3D Gaussian Splats

Vishnu Jaganathan, Hannah Hanyun Huang, Muhammad Zubair Irshad +3

Recently many techniques have emerged to create high quality 3D assets and scenes. When it comes to editing of these objects, however, existing approaches are either slow, compromi…

cs.HC20241 cited

Shaping Realities: Enhancing 3D Generative AI with Fabrication Constraints

Faraz Faruqi, Yingtao Tian, Vrushank Phadnis +2

Generative AI tools are becoming more prevalent in 3D modeling, enabling users to manipulate or create new models with text or images as inputs. This makes it easier for users to r…