◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Rameswar Panda

59 papers hereh-index 418.1k citations112 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author7
  • middle author40
  • last author8

Across the 55 of 59 papers where every author was matched, so the position is known.

fields
  • cs.CV37
  • cs.LG10
  • cs.CL6
  • cs.AI3
  • cs.DC1
  • cs.MM1
same name
  • Rameswar Panda — 7 papers, h 3
  • Rameswar Panda — 3 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20162025
most citedIA-RED2: Interpretability-Aware Redundancy Reduction for Vision Transformers

68 citations · 278 across the 31 of their papers we have counts for

collaborators
Showing 2023Show all

4 papers · 1 filter

cs.LG2023

Gated Linear Attention Transformers with Hardware-Efficient Training

Songlin Yang, Bailin Wang, Yikang Shen +2

Transformers with linear attention allow for efficient parallel training but can simultaneously be formulated as an RNN with 2D (matrix-valued) hidden states, thus enjoying linear-…

cs.CV2023★ 1 cited

Learning Human Action Recognition Representations Without Real Humans

Howard Zhong, Samarth Mishra, Donghyun Kim +7

Pre-training on massive video datasets has become essential to achieve high action recognition performance on smaller downstream datasets. However, most large-scale video datasets…

cs.CV2023

LangNav: Language as a Perceptual Representation for Navigation

Bowen Pan, Rameswar Panda, SouYoung Jin +4

We explore the use of language as a perceptual representation for vision-and-language navigation (VLN), with a focus on low-data settings. Our approach uses off-the-shelf vision sy…

cs.CV2023★ 12 cited

Dense and Aligned Captions (DAC) Promote Compositional Reasoning in VL Models

Sivan Doveh, Assaf Arbelle, Sivan Harary +9

Vision and Language (VL) models offer an effective method for aligning representation spaces of images and text, leading to numerous applications such as cross-modal retrieval, vis…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.