most citedWeakly-Supervised HOI Detection from Interaction Labels Only and Language/Vision-Language Priors

3 citations · 3 across the 7 of their papers we have counts for

collaborators

7 papers

cs.CV2024

Enhancing Weakly-Supervised Object Detection on Static Images through (Hallucinated) Motion

Cagri Gungor, Adriana Kovashka

While motion has garnered attention in various tasks, its potential as a modality for weakly-supervised object detection (WSOD) in static images remains unexplored. Our study intro…

cs.CV2024

Integrating Audio Narrations to Strengthen Domain Generalization in Multimodal First-Person Action Recognition

Cagri Gungor, Adriana Kovashka

First-person activity recognition is rapidly growing due to the widespread use of wearable cameras but faces challenges from domain shifts across different environments, such as va…

cs.CV2024

Incorporating Geo-Diverse Knowledge into Prompting for Increased Geographical Robustness in Object Recognition

Kyle Buettner, Sina Malakouti, Xiang Lorraine Li +1

Existing object recognition models have been shown to lack robustness in diverse geographical scenarios due to domain shifts in design and context. Class representations need to be…

cs.CV2023

Semi-Supervised Domain Generalization for Object Detection via Language-Guided Feature Alignment

Sina Malakouti, Adriana Kovashka

Existing domain adaptation (DA) and generalization (DG) methods in object detection enforce feature alignment in the visual space but face challenges like object appearance variabi…

cs.HC2023

Impact of Experiencing Misrecognition by Teachable Agents on Learning and Rapport

Yuya Asano, Diane Litman, Mingzhi Yu +4

While speech-enabled teachable agents have some advantages over typing-based ones, they are vulnerable to errors stemming from misrecognition by automatic speech recognition (ASR).…

cs.CV2023

Hypernymization of named entity-rich captions for grounding-based multi-modal pretraining

Giacomo Nebbia, Adriana Kovashka

Named entities are ubiquitous in text that naturally accompanies images, especially in domains such as news or Wikipedia articles. In previous work, named entities have been identi…