activity
20182022
most citedKnowledge Transfer via Distillation of Activation Boundaries Formed by Hidden Neurons

56 citations · 139 across the 7 of their papers we have counts for

collaborators
Showing cs.CVShow all

8 papers · 1 filter

cs.CV20225 cited

Group Generalized Mean Pooling for Vision Transformer

Byungsoo Ko, Han-Gyu Kim, Byeongho Heo +4

Vision Transformer (ViT) extracts the final representation from either class token or an average of all patch tokens, following the architecture of Transformer in Natural Language…

cs.CV202211 cited

An Extendable, Efficient and Effective Transformer-based Object Detector

Hwanjun Song, Deqing Sun, Sanghyuk Chun +5

Transformers have been widely used in numerous vision problems especially for visual recognition and detection. Detection transformers are the first fully end-to-end learning syste…

cs.CV20225 cited

Learning Features with Parameter-Free Layers

Dongyoon Han, YoungJoon Yoo, Beomyoung Kim +1

Trainable layers such as convolutional building blocks are the standard network design choices by learning parameters to capture the global context through successive spatial opera…

cs.CV2021

Rethinking Spatial Dimensions of Vision Transformers

Byeongho Heo, Sangdoo Yun, Dongyoon Han +3

Vision Transformer (ViT) extends the application range of transformers from language processing to computer vision tasks as being an alternative architecture against the existing c…

cs.CV2021

Re-labeling ImageNet: from Single to Multi-Labels, from Global to Localized Labels

Sangdoo Yun, Seong Joon Oh, Byeongho Heo +3

ImageNet has been arguably the most popular image classification benchmark, but it is also the one with a significant level of label noise. Recent studies have shown that many samp…

cs.CV202048 cited

VideoMix: Rethinking Data Augmentation for Video Classification

Sangdoo Yun, Seong Joon Oh, Byeongho Heo +2

State-of-the-art video action classifiers often suffer from overfitting. They tend to be biased towards specific objects and scene cues, rather than the foreground action content,…