3 papers
cs.CV2024
Hijacking Vision-and-Language Navigation Agents with Adversarial Environmental Attacks
Zijiao Yang, Xiangxi Shi, Eric Slyman +1
Assistive embodied agents that can be instructed in natural language to perform tasks in open-world environments have the potential to significantly impact labor tasks like manufac…
cs.CV2024
FairDeDup: Detecting and Mitigating Vision-Language Fairness Disparities in Semantic Dataset Deduplication
Eric Slyman, Stefan Lee, Scott Cohen +1
Recent dataset deduplication techniques have demonstrated that content-aware dataset pruning can dramatically reduce the cost of training Vision-Language Pretrained (VLP) models wi…
cs.LG2023
On the Behavior of Audio-Visual Fusion Architectures in Identity Verification Tasks
Daniel Claborne, Eric Slyman, Karl Pazdernik
We train an identity verification architecture and evaluate modifications to the part of the model that combines audio and visual representations, including in scenarios where one…