3 citations · 3 across the 7 of their papers we have counts for
7 papers
Enhancing Weakly-Supervised Object Detection on Static Images through (Hallucinated) Motion
Cagri Gungor, Adriana Kovashka
While motion has garnered attention in various tasks, its potential as a modality for weakly-supervised object detection (WSOD) in static images remains unexplored. Our study intro…
Integrating Audio Narrations to Strengthen Domain Generalization in Multimodal First-Person Action Recognition
Cagri Gungor, Adriana Kovashka
First-person activity recognition is rapidly growing due to the widespread use of wearable cameras but faces challenges from domain shifts across different environments, such as va…
Incorporating Geo-Diverse Knowledge into Prompting for Increased Geographical Robustness in Object Recognition
Kyle Buettner, Sina Malakouti, Xiang Lorraine Li +1
Existing object recognition models have been shown to lack robustness in diverse geographical scenarios due to domain shifts in design and context. Class representations need to be…
Semi-Supervised Domain Generalization for Object Detection via Language-Guided Feature Alignment
Sina Malakouti, Adriana Kovashka
Existing domain adaptation (DA) and generalization (DG) methods in object detection enforce feature alignment in the visual space but face challenges like object appearance variabi…
Impact of Experiencing Misrecognition by Teachable Agents on Learning and Rapport
Yuya Asano, Diane Litman, Mingzhi Yu +4
While speech-enabled teachable agents have some advantages over typing-based ones, they are vulnerable to errors stemming from misrecognition by automatic speech recognition (ASR).…
Hypernymization of named entity-rich captions for grounding-based multi-modal pretraining
Giacomo Nebbia, Adriana Kovashka
Named entities are ubiquitous in text that naturally accompanies images, especially in domains such as news or Wikipedia articles. In previous work, named entities have been identi…