3 papers
cs.LG2025
I-Diff: Structural Regularization for High-Fidelity Diffusion Models
Shakthi Perera, Dilum Fernando, H. L. P. Malshan +5
Denoising Diffusion Probabilistic Models (DDPMs) have significantly advanced generative AI, achieving impressive results in high-quality image and data generation. However, enhanci…
cs.CV2025
Stable Diffusion Models are Secretly Good at Visual In-Context Learning
Trevine Oorloff, Vishwanath Sindagi, Wele Gedara Chaminda Bandara +4
Large language models (LLM) in natural language processing (NLP) have demonstrated great potential for in-context learning (ICL) -- the ability to leverage a few sets of example pr…
cs.CV2025
:~Cataract Surgical Masked Autoencoder (MAE) based Pre-training
Nisarg A. Shah, Wele Gedara Chaminda Bandara, Shameema Skider +2
Automated analysis of surgical videos is crucial for improving surgical training, workflow optimization, and postoperative assessment. We introduce a CSMAE, Masked Autoencoder (MAE…