◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Mikhail Sirotenko

10 papers hereh-index 94.1k citations12 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author8
  • last author1

Across the 9 of 10 papers where every author was matched, so the position is known.

fields
  • cs.CV7
  • cs.LG3

identity via Semantic Scholar / OpenAlex

activity
20202025
most citedVideoPoet: A Large Language Model for Zero-Shot Video Generation

22 citations · 40 across the 9 of their papers we have counts for

collaborators
Showing 2025 · cs.LGShow all

2 papers · 2 filters

cs.LG2025

MINERVA: Evaluating Complex Video Reasoning

Arsha Nagrani, Sachit Menon, Ahmet Iscen +9

Multimodal LLMs are turning their focus to video benchmarks, however most video benchmarks only provide outcome supervision, with no intermediate or interpretable reasoning steps.…

cs.LG2025

Neptune: The Long Orbit to Benchmarking Long Video Understanding

Arsha Nagrani, Mingda Zhang, Ramin Mehran +10

We introduce Neptune, a benchmark for long video understanding that requires reasoning over long time horizons and across different modalities. Many existing video datasets and mod…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.