7 citations · 7 across the 6 of their papers we have counts for
6 papers · 1 filter
Open Ad-hoc Categorization with Contextualized Feature Learning
Zilin Wang, Sangwoo Mo, Stella X. Yu +2
Adaptive categorization of visual scenes is essential for AI agents to handle changing tasks. Unlike fixed common categories for plants or animals, ad-hoc categories are created dy…
VISTA: A Visual Analytics Framework to Enhance Foundation Model-Generated Data Labels
Xiwei Xuan, Xiaoqi Wang, Wenbin He +4
The advances in multi-modal foundation models (FMs) (e.g., CLIP and LLaVA) have facilitated the auto-labeling of large-scale datasets, enhancing model performance in challenging do…
USE: Universal Segment Embeddings for Open-Vocabulary Image Segmentation
Xiaoqi Wang, Wenbin He, Xiwei Xuan +8
The open-vocabulary image segmentation task involves partitioning images into semantically meaningful segments and classifying them with flexible text-defined categories. The recen…
A streamlined Approach to Multimodal Few-Shot Class Incremental Learning for Fine-Grained Datasets
Thang Doan, Sima Behpour, Xin Li +3
Few-shot Class-Incremental Learning (FSCIL) poses the challenge of retaining prior knowledge while learning from limited new data streams, all without overfitting. The rise of Visi…
AttributionScanner: A Visual Analytics System for Model Validation with Metadata-Free Slice Finding
Xiwei Xuan, Jorge Piazentin Ono, Liang Gou +2
Data slice finding is an emerging technique for validating machine learning (ML) models by identifying and analyzing subgroups in a dataset that exhibit poor performance, often cha…
Long-Distance Gesture Recognition using Dynamic Neural Networks
Shubhang Bhatnagar, Sharath Gopal, Narendra Ahuja +1
Gestures form an important medium of communication between humans and machines. An overwhelming majority of existing gesture recognition methods are tailored to a scenario where hu…