◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Siva Reddy

48 papers hereh-index 4611k citations104 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author26
  • last author20

Across the 47 of 48 papers where every author was matched, so the position is known.

fields
  • cs.CL35
  • cs.LG8
  • cs.CV4
  • cs.CY1
same name
  • Siva Reddy — 9 papers, h 2
  • Siva Reddy — 4 papers, h 2
  • Siva Reddy — 4 papers, h 2
  • Siva Reddy — 3 papers, h 3
  • Siva Reddy — 2 papers
  • Siva Reddy — 1 paper, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20162026
most citedMeDAL: Medical Abbreviation Disambiguation Dataset for Natural Language Understanding Pretraining

25 citations · 96 across the 19 of their papers we have counts for

collaborators
Showing cs.CVShow all

4 papers · 1 filter

cs.CV2026

CRAG-MM-Diagnostics: Enabling Stage-Wise Analysis of Knowledge-Intensive VQA

Hanseok Oh, Parishad BehnamGhader, Benno Krojer +4

Knowledge-Intensive Visual Question Answering (KI-VQA) benchmarks evaluate Vision-Language Models (VLMs) as multimodal knowledge assistants by requiring external information beyond…

cs.CV2024

Benchmarking Vision Language Models for Cultural Understanding

Shravan Nayak, Kanishk Jain, Rabiul Awal +5

Foundation models and vision-language pre-training have notably advanced Vision Language Models (VLMs), enabling multimodal processing of visual and linguistic data. However, their…

cs.CV2024

Learning Action and Reasoning-Centric Image Editing from Videos and Simulations

Benno Krojer, Dheeraj Vattikonda, Luis Lara +4

An image editing model should be able to perform diverse edits, ranging from object replacement, changing attributes or style, to performing actions or movement, which require many…

cs.CV2023

Are Diffusion Models Vision-And-Language Reasoners?

Benno Krojer, Elinor Poole-Dayan, Vikram Voleti +2

Text-conditioned image generation models have recently shown immense qualitative success using denoising diffusion processes. However, unlike discriminative vision-and-language mod…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.