◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Ankur Handa

6 papers hereh-index 6652 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3
  • last author2

Across the 5 of 6 papers where every author was matched, so the position is known.

fields
  • cs.CV3
  • cs.RO3
same name
  • Ankur Handa — 5 papers, h 33

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.CVShow all

3 papers · 1 filter

cs.CV2026

Cosmos 3: Omnimodal World Models for Physical AI

NVIDIA, :, Aditi +293

We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences within a unified mixture-of-t…

cs.CV2025

CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models

Qingqing Zhao, Yao Lu, Moo Jin Kim +12

Vision-language-action models (VLAs) have shown potential in leveraging pretrained vision-language models and diverse robot demonstrations for learning generalizable sensorimotor c…

cs.CV2024

Synthetica: Large Scale Synthetic Data for Robot Perception

Ritvik Singh, Jingzhou Liu, Karl Van Wyk +5

Vision-based object detectors are a crucial basis for robotics applications as they provide valuable information about object localisation in the environment. These need to ensure…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.