23 citations · 27 across the 3 of their papers we have counts for
5 papers · 1 filter
Efficient Visual Question Answering Pipeline for Autonomous Driving via Scene Region Compression
Yuliang Cai, Dongqiangzi Ye, Zitian Chen +1
Autonomous driving increasingly relies on Visual Question Answering (VQA) to enable vehicles to understand complex surroundings by analyzing visual inputs and textual queries. Curr…
Shot in the Dark: Few-Shot Learning with No Base-Class Labels
Zitian Chen, Subhransu Maji, Erik Learned-Miller
Few-shot learning aims to build classifiers for new classes from a small number of labeled examples and is commonly facilitated by access to examples from a distinct set of 'base c…
Cross-Supervised Object Detection
Zitian Chen, Zhiqiang Shen, Jiahui Yu +1
After learning a new object category from image-level annotations (with no object bounding boxes), humans are remarkably good at precisely localizing those objects. However, buildi…
Image Deformation Meta-Networks for One-Shot Learning
Zitian Chen, Yanwei Fu, Yu-Xiong Wang +3
Humans can robustly learn novel visual concepts even when images undergo various deformations and lose certain information. Mimicking the same behavior and synthesizing deformed in…
Multi-level Semantic Feature Augmentation for One-shot Learning
Zitian Chen, Yanwei Fu, Yinda Zhang +3
The ability to quickly recognize and learn new visual concepts from limited samples enables humans to swiftly adapt to new environments. This ability is enabled by semantic associa…