19 citations · 62 across the 14 of their papers we have counts for
24 papers
iQuery: Instruments as Queries for Audio-Visual Sound Separation
Jiaben Chen, Renrui Zhang, Dongze Lian +3
Current audio-visual separation methods share a standard architecture design where an audio encoder-decoder network is fused with visual encoding features at the encoder bottleneck…
HashEncoding: Autoencoding with Multiscale Coordinate Hashing
Lukas Zhornyak, Zhengjie Xu, Haoran Tang +1
We present HashEncoding, a novel autoencoding architecture that leverages a non-parametric multiscale coordinate hash function to facilitate a per-pixel decoder without convolution…
SoGCN: Second-Order Graph Convolutional Networks
Peihao Wang, Yuehao Wang, Hua Lin +1
Graph Convolutional Networks (GCN) with multi-hop aggregation is more expressive than one-hop GCN but suffers from higher model complexity. Finding the shortest aggregation range t…
Convolutional Ordinal Regression Forest for Image Ordinal Estimation
Haiping Zhu, Hongming Shan, Yuheng Zhang +5
Image ordinal estimation is to predict the ordinal label of a given image, which can be categorized as an ordinal regression problem. Recent methods formulate an ordinal regression…
Nested Scale Editing for Conditional Image Synthesis
Lingzhi Zhang, Jiancong Wang, Yinshuang Xu +4
We propose an image synthesis approach that provides stratified navigation in the latent code space. With a tiny amount of partial or very low-resolution image, our approach can co…
Potential Field: Interpretable and Unified Representation for Trajectory Prediction
Shan Su, Cheng Peng, Jianbo Shi +1
Predicting an agent's future trajectory is a challenging task given the complicated stimuli (environmental/inertial/social) of motion. Prior works learn individual stimulus from di…