collaborators

6 papers

cs.LG2025

Composite Classifier-Free Guidance for Multi-Modal Conditioning in Wind Dynamics Super-Resolution

Jacob Schnell, Aditya Makkar, Gunadi Gani +5

Various weather modelling problems (e.g., weather forecasting, optimizing turbine placements, etc.) require ample access to high-resolution, highly accurate wind data. Acquiring su…

cs.LG2025

SCALEX: Scalable Concept and Latent Exploration for Diffusion Models

E. Zhixuan Zeng, Yuhao Chen, Alexander Wong

Image generation models frequently encode social biases, including stereotypes tied to gender, race, and profession. Existing methods for analyzing these biases in diffusion models…

cs.AI2025

GraphPad: Inference-Time 3D Scene Graph Updates for Embodied Question Answering

Muhammad Qasim Ali, Saeejith Nair, Alexander Wong +2

Structured scene representations are a core component of embodied agents, helping to consolidate raw sensory streams into readable, modular, and searchable formats. Due to their hi…

cs.CV2025

SAMJAM: Zero-Shot Video Scene Graph Generation for Egocentric Kitchen Videos

Joshua Li, Fernando Jose Pena Cantu, Emily Yu +3

Video Scene Graph Generation (VidSGG) is an important topic in understanding dynamic kitchen environments. Current models for VidSGG require extensive training to produce scene gra…

cs.CV2024

MetaFood3D: 3D Food Dataset with Nutrition Values

Yuhao Chen, Jiangpeng He, Gautham Vinod +11

Food computing is both important and challenging in computer vision (CV). It significantly contributes to the development of CV algorithms due to its frequent presence in datasets…

cs.CL2024

Decoding Diffusion: A Scalable Framework for Unsupervised Analysis of Latent Space Biases and Representations Using Natural Language Prompts

E. Zhixuan Zeng, Yuhao Chen, Alexander Wong

Recent advances in image generation have made diffusion models powerful tools for creating high-quality images. However, their iterative denoising process makes understanding and i…