◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Y. Aytar

34 papers hereh-index 306.9k citations57 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author27
  • last author2

Across the 33 of 34 papers where every author was matched, so the position is known.

fields
  • cs.CV18
  • cs.LG7
  • cs.RO7
  • cs.CY1
  • cs.HC1

identity via Semantic Scholar / OpenAlex

activity
20162024
most citedSoundNet: Learning Sound Representations from Unlabeled Video

233 citations · 441 across the 25 of their papers we have counts for

collaborators
Showing 2023 · cs.CVShow all

3 papers · 2 filters

cs.CV2023

Learning from One Continuous Video Stream

João Carreira, Michael King, Viorica Pătrăucean +9

We introduce a framework for online learning from a single continuous video stream -- the way people and animals learn, without mini-batches, data augmentation or shuffling. This p…

cs.CV2023★ 5 cited

TAPIR: Tracking Any Point with per-frame Initialization and temporal Refinement

Carl Doersch, Yi Yang, Mel Vecerik +5

We present a novel model for Tracking Any Point (TAP) that effectively tracks any queried point on any physical surface throughout a video sequence. Our approach employs two stages…

cs.CV2023★ 12 cited

Perception Test: A Diagnostic Benchmark for Multimodal Video Models

Viorica Pătrăucean, Lucas Smaira, Ankush Gupta +21

We propose a novel multimodal video benchmark - the Perception Test - to evaluate the perception and reasoning skills of pre-trained multimodal models (e.g. Flamingo, SeViLA, or GP…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.