activity
20192026
most citedMaxViT: Multi-Axis Vision Transformer

28 citations · 82 across the 14 of their papers we have counts for

collaborators
Showing 2020Show all

6 papers · 1 filter

cs.CV2020

Multi-path Neural Networks for On-device Multi-domain Visual Classification

Qifei Wang, Junjie Ke, Joshua Greaves +9

Learning multiple domains/tasks with a single model is important for improving data efficiency and lowering inference cost for numerous vision tasks, especially on resource-constra…

eess.IV2020★ 4 cited

The Rate-Distortion-Accuracy Tradeoff: JPEG Case Study

Xiyang Luo, Hossein Talebi, Feng Yang +2

Handling digital images is almost always accompanied by a lossy compression in order to facilitate efficient transmission and storage. This introduces an unavoidable tension betwee…

eess.IV2020

GIFnets: Differentiable GIF Encoding Framework

Innfarn Yoo, Xiyang Luo, Yilin Wang +2

Graphics Interchange Format (GIF) is a widely used image file format. Due to the limited number of palette colors, GIF encoding often introduces color banding artifacts. Traditiona…

cs.CV2020★ 1 cited

Super-Resolving Commercial Satellite Imagery Using Realistic Training Data

Xiang Zhu, Hossein Talebi, Xinwei Shi +2

In machine learning based single image super-resolution, the degradation model is embedded in training data generation. However, most existing satellite image super-resolution meth…

eess.IV2020

Better Compression with Deep Pre-Editing

Hossein Talebi, Damien Kelly, Xiyang Luo +4

Could we compress images via standard codecs while avoiding visible artifacts? The answer is obvious -- this is doable as long as the bit budget is generous enough. What if the all…

cs.MM2020★ 5 cited

Distortion Agnostic Deep Watermarking

Xiyang Luo, Ruohan Zhan, Huiwen Chang +2

Watermarking is the process of embedding information into an image that can survive under distortions, while requiring the encoded image to have little or no perceptual difference…