51 citations · 364 across the 38 of their papers we have counts for
10 papers · 1 filter
LeRoP: A Learning-Based Modular Robot Photography Framework
Hao Kang, Jianming Zhang, Haoxiang Li +3
We introduce a novel framework for automatic capturing of human portraits. The framework allows the robot to follow a person to the desired location using a Person Re-identificatio…
Scaling Object Detection by Transferring Classification Weights
Jason Kuen, Federico Perazzi, Zhe Lin +2
Large scale object detection datasets are constantly increasing their size in terms of the number of classes and annotations count. Yet, the number of object-level categories annot…
Towards High-Resolution Salient Object Detection
Yi Zeng, Pingping Zhang, Jianming Zhang +2
Deep neural network based methods have made a significant breakthrough in salient object detection. However, they are typically limited to input images with low resolutions ($400\t…
Expressing Visual Relationships via Language
Hao Tan, Franck Dernoncourt, Zhe Lin +2
Describing images with text is a fundamental problem in vision-language research. Current studies in this domain mostly focus on single image captioning. However, in various real a…
Multitask Text-to-Visual Embedding with Titles and Clickthrough Data
Pranav Aggarwal, Zhe Lin, Baldo Faieta +1
Text-visual (or called semantic-visual) embedding is a central problem in vision-language research. It typically involves mapping of an image and a text description to a common fea…
Multimodal Style Transfer via Graph Cuts
Yulun Zhang, Chen Fang, Yilin Wang +4
An assumption widely used in recent neural style transfer methods is that image styles can be described by global statics of deep features like Gram or covariance matrices. Alterna…