5 papers
Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Gheorghe Comanici, Eric Bieber, Mike Schaekermann +3431
In this report, we introduce the Gemini 2.X model family: Gemini 2.5 Pro and Gemini 2.5 Flash, as well as our earlier Gemini 2.0 Flash and Flash-Lite models. Gemini 2.5 Pro is our…
Preserving Product Fidelity in Large Scale Image Recontextualization with Diffusion Models
Ishaan Malhi, Praneet Dutta, Ellie Talius +5
We present a framework for high-fidelity product image recontextualization using text-to-image diffusion models and a novel data augmentation pipeline. This pipeline leverages imag…
Scientific Discovery by Generating Counterfactuals using Image Translation
Arunachalam Narayanaswamy, Subhashini Venugopalan, Dale R. Webster +10
Model explanation techniques play a critical role in understanding the source of a model's performance and making its decisions transparent. Here we investigate if explanation tech…
It's easy to fool yourself: Case studies on identifying bias and confounding in bio-medical datasets
Subhashini Venugopalan, Arunachalam Narayanaswamy, Samuel Yang +10
Confounding variables are a well known source of nuisance in biomedical studies. They present an even greater challenge when we combine them with black-box machine learning techniq…
Predicting optical coherence tomography-derived diabetic macular edema grades from fundus photographs using deep learning
Avinash Varadarajan, Pinal Bavishi, Paisan Raumviboonsuk +15
Diabetic eye disease is one of the fastest growing causes of preventable blindness. With the advent of anti-VEGF (vascular endothelial growth factor) therapies, it has become incre…