108 citations · 132 across the 6 of their papers we have counts for
6 papers
PosterLayout: A New Benchmark and Approach for Content-aware Visual-Textual Presentation Layout
HsiaoYuan Hsu, Xiangteng He, Yuxin Peng +2
Content-aware visual-textual presentation layout aims at arranging spatial space on the given canvas for pre-defined elements, including text, logo, and underlay, which is a key to…
Multi-Behavior Recommendation with Cascading Graph Convolution Networks
Zhiyong Cheng, Sai Han, Fan Liu +3
Multi-behavior recommendation, which exploits auxiliary behaviors (e.g., click and cart) to help predict users' potential interactions on the target behavior (e.g., buy), is regard…
Enhancing Unsupervised Audio Representation Learning via Adversarial Sample Generation
Yulin Pan, Xiangteng He, Biao Gong +2
Existing audio analysis methods generally first transform the audio stream to spectrogram, and then feed it into CNN for further analysis. A standard CNN recognizes specific visual…
SIM-Trans: Structure Information Modeling Transformer for Fine-grained Visual Categorization
Hongbo Sun, Xiangteng He, Yuxin Peng
Fine-grained visual categorization (FGVC) aims at recognizing objects from similar subordinate categories, which is challenging and practical for human's accurate automatic recogni…
Team PKU-WICT-MIPL PIC Makeup Temporal Video Grounding Challenge 2022 Technical Report
Minghang Zheng, Dejie Yang, Zhongjie Ye +3
In this technical report, we briefly introduce the solutions of our team `PKU-WICT-MIPL' for the PIC Makeup Temporal Video Grounding (MTVG) Challenge in ACM-MM 2022. Given an untri…
The Application of Two-level Attention Models in Deep Convolutional Neural Network for Fine-grained Image Classification
Tianjun Xiao, Yichong Xu, Kuiyuan Yang +3
Fine-grained classification is challenging because categories can only be discriminated by subtle and local differences. Variances in the pose, scale or rotation usually make the p…