activity
20182022
most citedImproving Image Captioning by Leveraging Intra- and Inter-layer Global Representation in Transformer Network

18 citations · 25 across the 5 of their papers we have counts for

collaborators

12 papers

cs.IR20222 cited

EasyRec: An easy-to-use, extendable and efficient framework for building industrial recommendation systems

Mengli Cheng, Yue Gao, Guoqiang Liu +2

We present EasyRec, an easy-to-use, extendable and efficient recommendation framework for building industrial recommendation systems. Our EasyRec framework is superior in the follo…

cs.CV2022

Factored Attention and Embedding for Unstructured-view Topic-related Ultrasound Report Generation

Fuhai Chen, Rongrong Ji, Chengpeng Dai +4

Echocardiography is widely used to clinical practice for diagnosis and treatment, e.g., on the common congenital heart defects. The traditional manual manipulation is error-prone d…

cs.CV2021

View-Guided Point Cloud Completion

Xuancheng Zhang, Yutong Feng, Siqi Li +5

This paper presents a view-guided solution for the task of point cloud completion. Unlike most existing methods directly inferring the missing points using shape priors, we address…

cs.LG2021

ReCU: Reviving the Dead Weights in Binary Neural Networks

Zihan Xu, Mingbao Lin, Jianzhuang Liu +5

Binary neural networks (BNNs) have received increasing attention due to their superior reductions of computation and memory. Most existing works focus on either lessening the quant…

cs.CV202018 cited

Improving Image Captioning by Leveraging Intra- and Inter-layer Global Representation in Transformer Network

Jiayi Ji, Yunpeng Luo, Xiaoshuai Sun +5

Transformer-based architectures have shown great success in image captioning, where object regions are encoded and then attended into the vectorial representations to guide the cap…

cs.CV20205 cited

Attention-based Multi-modal Fusion Network for Semantic Scene Completion

Siqi Li, Changqing Zou, Yipeng Li +2

This paper presents an end-to-end 3D convolutional network named attention-based multi-modal fusion network (AMFNet) for the semantic scene completion (SSC) task of inferring the o…