activity
20162026
most citedAdan: Adaptive Nesterov Momentum Algorithm for Faster Optimizing Deep Models

63 citations · 504 across the 62 of their papers we have counts for

collaborators
Showing 2021 · cs.LGShow all

7 papers · 2 filters

cs.LG2021★ 13 cited

On Training Implicit Models

Zhengyang Geng, Xin-Yu Zhang, Shaojie Bai +2

This paper focuses on training implicit models of infinite layers. Specifically, previous works employ implicit differentiation and solve the exact gradient for the backward propag…

cs.LG2021

Pareto Adversarial Robustness: Balancing Spatial Robustness and Sensitivity-based Robustness

Ke Sun, Mingjie Li, Zhouchen Lin

Adversarial robustness, which primarily comprises sensitivity-based robustness and spatial robustness, plays an integral part in achieving robust generalization. In this paper, we…

cs.LG2021★ 3 cited

Residual Relaxation for Multi-view Representation Learning

Yifei Wang, Zhengyang Geng, Feng Jiang +4

Multi-view methods learn representations by aligning multiple views of the same image and their performance largely depends on the choice of data augmentation. In this paper, we no…

cs.LG2021★ 1 cited

Leveraged Weighted Loss for Partial Label Learning

Hongwei Wen, Jingyi Cui, Hanyuan Hang +3

As an important branch of weakly supervised learning, partial label learning deals with data where each instance is assigned with a set of candidate labels, whereas only one of the…

cs.LG2021★ 2 cited

Optimization Induced Equilibrium Networks

Xingyu Xie, Qiuhao Wang, Zenan Ling +4

Implicit equilibrium models, i.e., deep neural networks (DNNs) defined by implicit equations, have been becoming more and more attractive recently. In this paper, we investigate an…

cs.LG2021

Dissecting the Diffusion Process in Linear Graph Convolutional Networks

Yifei Wang, Yisen Wang, Jiansheng Yang +1

Graph Convolutional Networks (GCNs) have attracted more and more attentions in recent years. A typical GCN layer consists of a linear feature propagation step and a nonlinear trans…