papers

Publications (26)

cs.CV2025

A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation

Andrew Z. Wang, Songwei Ge, Tero Karras +2

cs.CV2024

Coherent Zero-Shot Visual Instruction Generation

Quynh Phung, Songwei Ge, Jia-Bin Huang

cs.CV2024

Preserve Your Own Correlation: A Noise Prior for Video Diffusion Models

Songwei Ge, Seungjun Nah, Guilin Liu +7

cs.LG2019

Getting Topology and Point Cloud Generation to Mesh

Austin Dill, Chun-Liang Li, Songwei Ge +1

cs.CV2022

Hyperbolic Contrastive Learning for Visual Representations beyond Objects

Songwei Ge, Shlok Mishra, Simon Kornblith +2

cs.AI2018

Hallucinating Point Cloud into 3D Sculptural Object

Chun-Liang Li, Eunsu Kang, Songwei Ge +4

cs.CV2021

Creative Sketch Generation

Songwei Ge, Vedanuj Goswami, C. Lawrence Zitnick +1

cs.CV2019

Learning Robust Global Representations by Penalizing Local Predictive Power

Haohan Wang, Songwei Ge, Eric P. Xing +1

cs.CV2025

PAD3R: Pose-Aware Dynamic 3D Reconstruction from Casual Videos

Ting-Hsuan Liao, Haowen Liu, Yiran Xu +3

cs.CV2025

Expressive Text-to-Image Generation with Rich Text

Songwei Ge, Taesung Park, Jun-Yan Zhu +1

cs.CV2024

Rethinking Score Distillation as a Bridge Between Image Distributions

David McAllister, Songwei Ge, Jia-Bin Huang +4

cs.CV2025

Cosmos World Foundation Model Platform for Physical AI

NVIDIA, :, Niket Agarwal +76

cs.CV2025

Illusion3D: 3D Multiview Illusion with 2D Diffusion Priors

Yue Feng, Vaibhav Sanjay, Spencer Lutz +3

cs.LG2021

Shift Invariance Can Reduce Adversarial Robustness

Songwei Ge, Vasu Singla, Ronen Basri +1

cs.LG2025

Flow Matching Policy Gradients

David McAllister, Songwei Ge, Brent Yi +5

cs.CV2023

Text-driven Visual Synthesis with Latent Diffusion Prior

Ting-Hsuan Liao, Songwei Ge, Yiran Xu +3

cs.GR2020

Learned Interpolation for 3D Generation

Austin Dill, Songwei Ge, Eunsu Kang +2

cs.IR2019

From Text to Sound: A Preliminary Study on Retrieving Sound Effects to Radio Stories

Songwei Ge, Curtis Xuan, Ruihua Song +3

cs.LG2019

Developing Creative AI to Generate Sculptural Objects

Songwei Ge, Austin Dill, Eunsu Kang +4

cs.IR2019

Personalizing Search Results Using Hierarchical RNN with Query-aware Attention

Songwei Ge, Zhicheng Dou, Zhengbao Jiang +2

cs.CV2022

MUGEN: A Playground for Video-Audio-Text Multimodal Understanding and GENeration

Thomas Hayes, Songyang Zhang, Xi Yin +6

cs.CV2024

On the Content Bias in Fréchet Video Distance

Songwei Ge, Aniruddha Mahapatra, Gaurav Parmar +2

cs.CV2022

Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Songwei Ge, Thomas Hayes, Harry Yang +5

cs.CL2021

Visual Conceptual Blending with Large-scale Language and Vision Models

Songwei Ge, Devi Parikh

cs.CV2022

Robust Contrastive Learning Using Negative Samples with Diminished Semantics

Songwei Ge, Shlok Mishra, Haohan Wang +2

cs.CV2023

Grounded Text-to-Image Synthesis with Attention Refocusing

Quynh Phung, Songwei Ge, Jia-Bin Huang