6 papers
Composite Classifier-Free Guidance for Multi-Modal Conditioning in Wind Dynamics Super-Resolution
Jacob Schnell, Aditya Makkar, Gunadi Gani +5
Various weather modelling problems (e.g., weather forecasting, optimizing turbine placements, etc.) require ample access to high-resolution, highly accurate wind data. Acquiring su…
SCALEX: Scalable Concept and Latent Exploration for Diffusion Models
E. Zhixuan Zeng, Yuhao Chen, Alexander Wong
Image generation models frequently encode social biases, including stereotypes tied to gender, race, and profession. Existing methods for analyzing these biases in diffusion models…
GraphPad: Inference-Time 3D Scene Graph Updates for Embodied Question Answering
Muhammad Qasim Ali, Saeejith Nair, Alexander Wong +2
Structured scene representations are a core component of embodied agents, helping to consolidate raw sensory streams into readable, modular, and searchable formats. Due to their hi…
SAMJAM: Zero-Shot Video Scene Graph Generation for Egocentric Kitchen Videos
Joshua Li, Fernando Jose Pena Cantu, Emily Yu +3
Video Scene Graph Generation (VidSGG) is an important topic in understanding dynamic kitchen environments. Current models for VidSGG require extensive training to produce scene gra…
MetaFood3D: 3D Food Dataset with Nutrition Values
Yuhao Chen, Jiangpeng He, Gautham Vinod +11
Food computing is both important and challenging in computer vision (CV). It significantly contributes to the development of CV algorithms due to its frequent presence in datasets…
Decoding Diffusion: A Scalable Framework for Unsupervised Analysis of Latent Space Biases and Representations Using Natural Language Prompts
E. Zhixuan Zeng, Yuhao Chen, Alexander Wong
Recent advances in image generation have made diffusion models powerful tools for creating high-quality images. However, their iterative denoising process makes understanding and i…