67 citations · 116 across the 7 of their papers we have counts for
11 papers · 1 filter
Argus++: Robust Real-time Activity Detection for Unconstrained Video Streams with Overlapping Cube Proposals
Lijun Yu, Yijun Qian, Wenhe Liu +1
Activity detection is one of the attractive computer vision tasks to exploit the video streams captured by widely installed cameras. Although achieving impressive performance, conv…
Subspace Representation Learning for Few-shot Image Classification
Ting-Yao Hu, Zhi-Qi Cheng, Alexander G. Hauptmann
In this paper, we propose a subspace representation learning (SRL) framework to tackle few-shot image classification tasks. It exploits a subspace in local CNN feature space to rep…
Person Search Challenges and Solutions: A Survey
Xiangtan Lin, Pengzhen Ren, Yun Xiao +2
Person search has drawn increasing attention due to its real-world applications and research significance. Person search aims to find a probe person in a gallery of scene images wi…
Multilingual Multimodal Pre-training for Zero-Shot Cross-Lingual Transfer of Vision-Language Models
Po-Yao Huang, Mandela Patrick, Junjie Hu +3
This paper studies zero-shot cross-lingual transfer of vision-language models. Specifically, we focus on multilingual text-to-video search and propose a Transformer-based model tha…
Pose Guided Person Image Generation with Hidden p-Norm Regression
Ting-Yao Hu, Alexander G. Hauptmann
In this paper, we propose a novel approach to solve the pose guided person image generation task. We assume that the relation between pose and appearance information can be describ…
Pixel-Level Cycle Association: A New Perspective for Domain Adaptive Semantic Segmentation
Guoliang Kang, Yunchao Wei, Yi Yang +2
Domain adaptive semantic segmentation aims to train a model performing satisfactory pixel-level predictions on the target with only out-of-domain (source) annotations. The conventi…