35 citations · 40 across the 11 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
SurGen: Text-Guided Diffusion Model for Surgical Video Generation
Joseph Cho, Samuel Schmidgall, Cyril Zakka +4
Diffusion-based video generation models have made significant strides, producing outputs with improved visual fidelity, temporal coherence, and user control. These advancements hol…
cs.CV2024
GP-VLS: A general-purpose vision language model for surgery
Samuel Schmidgall, Joseph Cho, Cyril Zakka +1
Surgery requires comprehensive medical knowledge, visual assessment skills, and procedural expertise. While recent surgical AI models have focused on solving task-specific problems…