activity
20162026
most citedMVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images

167 citations · 443 across the 80 of their papers we have counts for

collaborators
Showing 2022Show all

7 papers · 1 filter

cs.CV2022★ 3 cited

BiViT: Extremely Compressed Binary Vision Transformer

Yefei He, Zhenyu Lou, Luoming Zhang +4

Model binarization can significantly compress model size, reduce energy consumption, and accelerate inference through efficient bit-wise operations. Although binarizing convolution…

cs.CV2022★ 16 cited

EcoFormer: Energy-Saving Attention with Linear Complexity

Jing Liu, Zizheng Pan, Haoyu He +2

Transformer is a transformative framework that models sequential data and has achieved remarkable performance on a wide range of tasks, but with high computational and energy cost.…

cs.CV2022★ 3 cited

FocusFormer: Focusing on What We Need via Architecture Sampler

Jing Liu, Jianfei Cai, Bohan Zhuang

Vision Transformers (ViTs) have underpinned the recent breakthroughs in computer vision. However, designing the architectures of ViTs is laborious and heavily relies on expert know…

cs.CV2022★ 1 cited

An Efficient Spatio-Temporal Pyramid Transformer for Action Detection

Yuetian Weng, Zizheng Pan, Mingfei Han +2

The task of action detection aims at deducing both the action category and localization of the start and end moment for each action instance in a long, untrimmed video. While visio…

cs.CV2022★ 116 cited

Fast Vision Transformers with HiLo Attention

Zizheng Pan, Jianfei Cai, Bohan Zhuang

Vision Transformers (ViTs) have triggered the most recent and significant breakthroughs in computer vision. Their efficient designs are mostly guided by the indirect metric of comp…

cs.CV2022

Dynamic Focus-aware Positional Queries for Semantic Segmentation

Haoyu He, Jianfei Cai, Zizheng Pan +4

The DETR-like segmentors have underpinned the most recent breakthroughs in semantic segmentation, which end-to-end train a set of queries representing the class prototypes or targe…