11 citations · 12 across the 3 of their papers we have counts for
3 papers
INQUIRE: A Natural World Text-to-Image Retrieval Benchmark
Edward Vendrow, Omiros Pantazis, Alexander Shepard +5
We introduce INQUIRE, a text-to-image retrieval benchmark designed to challenge multimodal vision-language models on expert-level queries. INQUIRE includes iNaturalist 2024 (iNat24…
Agile Modeling: From Concept to Classifier in Minutes
Otilia Stretcu, Edward Vendrow, Kenji Hata +15
The application of computer vision to nuanced subjective use cases is growing. While crowdsourcing has served the vision community well for most objective tasks (such as labeling a…
SoMoFormer: Multi-Person Pose Forecasting with Transformers
Edward Vendrow, Satyajit Kumar, Ehsan Adeli +1
Human pose forecasting is a challenging problem involving complex human body motion and posture dynamics. In cases that there are multiple people in the environment, one's motion m…