◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

M. Donoser

3 papers hereh-index 232.4k citations78 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author1

Across the 2 of 3 papers where every author was matched, so the position is known.

fields
  • cs.CV3

identity via Semantic Scholar / OpenAlex

activity
20212023
most citedContrastive Language-Action Pre-training for Temporal Localization

5 citations · 9 across the 3 of their papers we have counts for

collaborators
Showing cs.CVShow all

3 papers · 1 filter

cs.CV2023★ 3 cited

iEdit: Localised Text-guided Image Editing with Weak Supervision

Rumeysa Bodur, Erhan Gundogdu, Binod Bhattarai +3

Diffusion models (DMs) can generate realistic images with text guidance using large-scale datasets. However, they demonstrate limited controllability in the output space of the gen…

cs.CV2022★ 5 cited

Contrastive Language-Action Pre-training for Temporal Localization

Mengmeng Xu, Erhan Gundogdu, Maksim Lapin +3

Long-form video understanding requires designing approaches that are able to temporally localize activities or language. End-to-end training for such tasks is limited by the comput…

cs.CV2021★ 1 cited

Revamping Cross-Modal Recipe Retrieval with Hierarchical Transformers and Self-supervised Learning

Amaia Salvador, Erhan Gundogdu, Loris Bazzani +1

Cross-modal recipe retrieval has recently gained substantial attention due to the importance of food in people's lives, as well as the availability of vast amounts of digital cooki…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.