EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks
arXiv:1905.11946
Abstract
Convolutional Neural Networks (ConvNets) are commonly developed at a fixed resource budget, and then scaled up for better accuracy if more resources are available. In this paper, we systematically study model scaling and identify that carefully balancing network depth, width, and resolution can lead to better performance. Based on this observation, we propose a new scaling method that uniformly scales all dimensions of depth/width/resolution using a simple yet highly effective compound coefficient. We demonstrate the effectiveness of this method on scaling up MobileNets and ResNet. To go even further, we use neural architecture search to design a new baseline network and scale it up to obtain a family of models, called EfficientNets, which achieve much better accuracy and efficiency than previous ConvNets. In particular, our EfficientNet-B7 achieves state-of-the-art 84.3% top-1 accuracy on ImageNet, while being 8.4x smaller and 6.1x faster on inference than the best existing ConvNet. Our EfficientNets also transfer well and achieve state-of-the-art accuracy on CIFAR-100 (91.7%), Flowers (98.8%), and 3 other transfer learning datasets, with an order of magnitude fewer parameters. Source code is at https://github.com/tensorflow/tpu/tree/master/models/official/efficientnet.
ICML 2019
References in corpus (3)
Cited by in corpus (398)
- Learning Transferable Visual Models From Natural Language Supervision
- Knowledge Distillation: A Survey
- A Comprehensive Review of YOLO Architectures in Computer Vision: From YOLOv1 to YOLOv8 and YOLO-NAS
- U-Net and its variants for medical image segmentation: theory and applications
- Aleatoric and Epistemic Uncertainty in Machine Learning: An Introduction to Concepts and Methods
- Ensemble Distillation for Robust Model Fusion in Federated Learning
- Hyper-Parameter Optimization: A Review of Algorithms and Applications
- Deep Industrial Image Anomaly Detection: A Survey
- Avoiding Overfitting: A Survey on Regularization Methods for Convolutional Neural Networks
- Real-Time Polyp Detection, Localization and Segmentation in Colonoscopy Using Deep Learning
- Visual Transformers: Token-based Image Representation and Processing for Computer Vision
- Sustainable AI: Environmental Implications, Challenges and Opportunities
- Waste detection in Pomerania: non-profit project for detecting waste in environment
- Deep Gradient Learning for Efficient Camouflaged Object Detection
- LocalViT: Analyzing Locality in Vision Transformers
- A Primer on Motion Capture with Deep Learning: Principles, Pitfalls and Perspectives
- Modality specific U-Net variants for biomedical image segmentation: A survey
- An Overview of Deep Semi-Supervised Learning
- Deepfake Detection by Human Crowds, Machines, and Machine-informed Crowds
- DepthFormer: Exploiting Long-Range Correlation and Local Information for Accurate Monocular Depth Estimation
- Integer Quantization for Deep Learning Inference: Principles and Empirical Evaluation
- Revisiting ResNets: Improved Training and Scaling Strategies
- Quantization and Deployment of Deep Neural Networks on Microcontrollers
- A Comprehensive Survey of Neural Architecture Search: Challenges and Solutions
- GhostNets on Heterogeneous Devices via Cheap Operations
- HeteroFL: Computation and Communication Efficient Federated Learning for Heterogeneous Clients
- Group Knowledge Transfer: Federated Learning of Large CNNs at the Edge
- On Empirical Comparisons of Optimizers for Deep Learning
- DeepSOCIAL: Social Distancing Monitoring and Infection Risk Assessment in COVID-19 Pandemic
- Neonatal seizure detection from raw multi-channel EEG using a fully convolutional architecture
- Self-supervised Pretraining of Visual Features in the Wild
- Advances in Deep Concealed Scene Understanding
- Optimization for deep learning: theory and algorithms
- End2End Occluded Face Recognition by Masking Corrupted Features
- Robust Multi-Task Learning and Online Refinement for Spacecraft Pose Estimation across Domain Gap
- Combining a Convolutional Neural Network with Autoencoders to Predict the Survival Chance of COVID-19 Patients
- Machine learning and AI-based approaches for bioactive ligand discovery and GPCR-ligand recognition
- The Effects of Skin Lesion Segmentation on the Performance of Dermatoscopic Image Classification
- Adaptive Inference through Early-Exit Networks: Design, Challenges and Directions
- Sharpness-Aware Minimization for Efficiently Improving Generalization
- DeepCorn: A Semi-Supervised Deep Learning Method for High-Throughput Image-Based Corn Kernel Counting and Yield Estimation
- PP-LCNet: A Lightweight CPU Convolutional Neural Network
- Optimized Deep Encoder-Decoder Methods for Crack Segmentation
- LeViT: a Vision Transformer in ConvNet's Clothing for Faster Inference
- Segmentation of cell-level anomalies in electroluminescence images of photovoltaic modules
- Deep Metric Learning-based Image Retrieval System for Chest Radiograph and its Clinical Applications in COVID-19
- MEDIC: A Multi-Task Learning Dataset for Disaster Image Classification
- Discovering Parametric Activation Functions
- FIgLib & SmokeyNet: Dataset and Deep Learning Model for Real-Time Wildland Fire Smoke Detection
- AdamP: Slowing Down the Slowdown for Momentum Optimizers on Scale-invariant Weights
- Local Motion Planner for Autonomous Navigation in Vineyards with a RGB-D Camera-Based Algorithm and Deep Learning Synergy
- InsPLAD: A Dataset and Benchmark for Power Line Asset Inspection in UAV Images
- SEED: Self-supervised Distillation For Visual Representation
- NBDT: Neural-Backed Decision Trees
- Fixing the train-test resolution discrepancy: FixEfficientNet
- RepMLP: Re-parameterizing Convolutions into Fully-connected Layers for Image Recognition
- Asymmetric Loss For Multi-Label Classification
- Batch Normalization Biases Residual Blocks Towards the Identity Function in Deep Networks
- ResKD: Residual-Guided Knowledge Distillation
- Clover: Toward Sustainable AI with Carbon-Aware Machine Learning Inference Service
- WHENet: Real-time Fine-Grained Estimation for Wide Range Head Pose
- r/Fakeddit: A New Multimodal Benchmark Dataset for Fine-grained Fake News Detection
- SEMI-FND: Stacked Ensemble Based Multimodal Inference For Faster Fake News Detection
- Multiscale Vision Transformers
- Deep Learning for Distinguishing Normal versus Abnormal Chest Radiographs and Generalization to Unseen Diseases
- Panoptic Segmentation Meets Remote Sensing
- Compacting Deep Neural Networks for Internet of Things: Methods and Applications
- HEROHE Challenge: assessing HER2 status in breast cancer without immunohistochemistry or in situ hybridization
- Affinity and Diversity: Quantifying Mechanisms of Data Augmentation
- Rotate to Attend: Convolutional Triplet Attention Module
- Medical Imaging and Machine Learning
- Advancing Additive Manufacturing through Deep Learning: A Comprehensive Review of Current Progress and Future Challenges
- Introduction to Camera Pose Estimation with Deep Learning
- RandStainNA: Learning Stain-Agnostic Features from Histology Slides by Bridging Stain Augmentation and Normalization
- The DeepFake Detection Challenge (DFDC) Dataset
- Trends in Integration of Vision and Language Research: A Survey of Tasks, Datasets, and Methods
- QKD: Quantization-aware Knowledge Distillation
- Non-imaging single-pixel sensing with optimized binary modulation
- Identifying Melanoma Images using EfficientNet Ensemble: Winning Solution to the SIIM-ISIC Melanoma Classification Challenge
- Semi-Supervised Neural Architecture Search
- RepVGG: Making VGG-style ConvNets Great Again
- Detecting Visual Design Principles in Art and Architecture through Deep Convolutional Neural Networks
- Systematically Measuring Ultra-Diffuse Galaxies (SMUDGes). II. Expanded Survey Description and the Stripe 82 Catalog
- Learning Physical Graph Representations from Visual Scenes
- MobileDets: Searching for Object Detection Architectures for Mobile Accelerators
- Refiner: Refining Self-attention for Vision Transformers
- Deep learning for the detection of machining vibration chatter
- Danish Fungi 2020 -- Not Just Another Image Recognition Dataset
- MKANet: A Lightweight Network with Sobel Boundary Loss for Efficient Land-cover Classification of Satellite Remote Sensing Imagery
- Rethinking BiSeNet For Real-time Semantic Segmentation
- QUCoughScope: An Artificially Intelligent Mobile Application to Detect Asymptomatic COVID-19 Patients using Cough and Breathing Sounds
- Real-Time Human Pose Estimation on a Smart Walker using Convolutional Neural Networks
- Contrastive Self-supervised Neural Architecture Search
- Medical Image Classification Using Transfer Learning and Chaos Game Optimization on the Internet of Medical Things
- HS-ResNet: Hierarchical-Split Block on Convolutional Neural Network
- An original framework for Wheat Head Detection using Deep, Semi-supervised and Ensemble Learning within Global Wheat Head Detection (GWHD) Dataset
- Sample-Efficient Neural Architecture Search by Learning Action Space
- ResNet strikes back: An improved training procedure in timm
- Products-10K: A Large-scale Product Recognition Dataset
- An embedded system for the automated generation of labeled plant images to enable machine learning applications in agriculture
- Size Matters
- Forensic Dental Age Estimation Using Modified Deep Learning Neural Network
- FBNetV2: Differentiable Neural Architecture Search for Spatial and Channel Dimensions
- A Study of the Generalizability of Self-Supervised Representations
- Outside the Box: Abstraction-Based Monitoring of Neural Networks
- One Proxy Device Is Enough for Hardware-Aware Neural Architecture Search
- RatLesNetv2: A Fully Convolutional Network for Rodent Brain Lesion Segmentation
- UR2KiD: Unifying Retrieval, Keypoint Detection, and Keypoint Description without Local Correspondence Supervision
- The Implicit and Explicit Regularization Effects of Dropout
- DSNAS: Direct Neural Architecture Search without Parameter Retraining
- 2D bidirectional gated recurrent unit convolutional Neural networks for end-to-end violence detection In videos
- Automated Detection of Cat Facial Landmarks
- Preferences Prediction using a Gallery of Mobile Device based on Scene Recognition and Object Detection
- Learning image representations for anomaly detection: application to discovery of histological alterations in drug development
- SpecDETR: A transformer-based hyperspectral point object detection network
- Predictive Analysis of Diabetic Retinopathy with Transfer Learning
- Vision Permutator: A Permutable MLP-Like Architecture for Visual Recognition
- Adaptive Linear Span Network for Object Skeleton Detection
- Learned Threshold Pruning
- Phase Recognition in Contrast-Enhanced CT Scans based on Deep Learning and Random Sampling
- Cervical Optical Coherence Tomography Image Classification Based on Contrastive Self-Supervised Texture Learning
- Fixing the train-test resolution discrepancy
- Distillation-based fabric anomaly detection
- Optimized Three Deep Learning Models Based-PSO Hyperparameters for Beijing PM2.5 Prediction
- Generalizable multi-task, multi-domain deep segmentation of sparse pediatric imaging datasets via multi-scale contrastive regularization and multi-joint anatomical priors
- Giving Commands to a Self-Driving Car: How to Deal with Uncertain Situations?
- AdaFuse: Adaptive Temporal Fusion Network for Efficient Action Recognition
- Human Gender Prediction Based on Deep Transfer Learning from Panoramic Radiograph Images
- Ekya: Continuous Learning of Video Analytics Models on Edge Compute Servers
- Squeeze-and-Attention Networks for Semantic Segmentation
- DFUC2020: Analysis Towards Diabetic Foot Ulcer Detection
- Progressive DARTS: Bridging the Optimization Gap for NAS in the Wild
- Efficient, high-performance pancreatic segmentation using multi-scale feature extraction
- Open-Vocabulary Animal Keypoint Detection with Semantic-feature Matching
- DISCOVER: 2-D Multiview Summarization of Optical Coherence Tomography Angiography for Automatic Diabetic Retinopathy Diagnosis
- Efficient Visual Recognition with Deep Neural Networks: A Survey on Recent Advances and New Directions
- HAWQV3: Dyadic Neural Network Quantization
- Deep4Air: A Novel Deep Learning Framework for Airport Airside Surveillance
- Finite size corrections for neural network Gaussian processes
- Towards Assessing the Synthetic-to-Measured Adversarial Vulnerability of SAR ATR
- MoViNets: Mobile Video Networks for Efficient Video Recognition
- Pollen Grain Microscopic Image Classification Using an Ensemble of Fine-Tuned Deep Convolutional Neural Networks
- CoolMomentum: A Method for Stochastic Optimization by Langevin Dynamics with Simulated Annealing
- MSD: Multi-Self-Distillation Learning via Multi-classifiers within Deep Neural Networks
- Lightweight Regression Model with Prediction Interval Estimation for Computer Vision-based Winter Road Surface Condition Monitoring
- Performance of GAN-based augmentation for deep learning COVID-19 image classification
- An interpretable classifier for high-resolution breast cancer screening images utilizing weakly supervised localization
- LSQ+: Improving low-bit quantization through learnable offsets and better initialization
- On the Predictability of Pruning Across Scales
- Temporal-Coded Deep Spiking Neural Network with Easy Training and Robust Performance
- HandAugment: A Simple Data Augmentation Method for Depth-Based 3D Hand Pose Estimation
- I-BERT: Integer-only BERT Quantization
- Post-Training Piecewise Linear Quantization for Deep Neural Networks
- Artificial mental phenomena: Psychophysics as a framework to detect perception biases in AI models
- DNA: Differentiable Network-Accelerator Co-Search
- Exploring the Efficacy of Base Data Augmentation Methods in Deep Learning-Based Radiograph Classification of Knee Joint Osteoarthritis
- Deep residential representations: Using unsupervised learning to unlock elevation data for geo-demographic prediction
- Neural Transformers for Intraductal Papillary Mucosal Neoplasms (IPMN) Classification in MRI images
- ASFD: Automatic and Scalable Face Detector
- A Computer Vision-Based Approach for Driver Distraction Recognition using Deep Learning and Genetic Algorithm Based Ensemble
- Efficient Differentiable Neural Architecture Search with Meta Kernels
- SM-NAS: Structural-to-Modular Neural Architecture Search for Object Detection
- Deep Learning-based automated classification of Chinese Speech Sound Disorders
- Rethinking Channel Dimensions for Efficient Model Design
- Segmentation-Free Outcome Prediction from Head and Neck Cancer PET/CT Images: Deep Learning-Based Feature Extraction from Multi-Angle Maximum Intensity Projections (MA-MIPs)
- Byzantine-Robust and Privacy-Preserving Framework for FedML
- Automatic Diagnosis of Pneumothorax from Chest Radiographs: A Systematic Literature Review
- AdversarialNAS: Adversarial Neural Architecture Search for GANs
- Fair Comparison: Quantifying Variance in Resultsfor Fine-grained Visual Categorization
- Control Distance IoU and Control Distance IoU Loss Function for Better Bounding Box Regression
- ROI-Aware Multiscale Cross-Attention Vision Transformer for Pest Image Identification
- Flow imaging as an alternative to pressure transducers through vision transformers and convolutional neural networks
- AttentiveNAS: Improving Neural Architecture Search via Attentive Sampling
- Highly Efficient Salient Object Detection with 100K Parameters
- Time for a Background Check! Uncovering the impact of Background Features on Deep Neural Networks
- BETANAS: BalancEd TrAining and selective drop for Neural Architecture Search
- Multiresolution Convolutional Autoencoders
- Beyond the Eye: A Relational Model for Early Dementia Detection Using Retinal OCTA Images
- Automatic cerebral hemisphere segmentation in rat MRI with lesions via attention-based convolutional neural networks
- MixMatch Domain Adaptaion: Prize-winning solution for both tracks of VisDA 2019 challenge
- Rethinking FUN: Frequency-Domain Utilization Networks
- Deep neural networks approach to microbial colony detection -- a comparative analysis
- Automatic Group Cohesiveness Detection With Multi-modal Features
- Neural Semi-supervised Learning for Text Classification Under Large-Scale Pretraining
- Effective and Efficient Computation with Multiple-timescale Spiking Recurrent Neural Networks
- Cyclic Differentiable Architecture Search
- Learning on Hardware: A Tutorial on Neural Network Accelerators and Co-Processors
- Activate or Not: Learning Customized Activation
- Any-Precision Deep Neural Networks
- Latent Causal Invariant Model
- Hierarchical Neural Architecture Search via Operator Clustering
- LeYOLO, New Embedded Architecture for Object Detection
- MnasFPN: Learning Latency-aware Pyramid Architecture for Object Detection on Mobile Devices
- Recognition-Aware Learned Image Compression
- Improving One-shot NAS by Suppressing the Posterior Fading
- SparseRT: Accelerating Unstructured Sparsity on GPUs for Deep Learning Inference
- Structured Convolutions for Efficient Neural Network Design
- Graph-based Interpolation of Feature Vectors for Accurate Few-Shot Classification
- Semi-Supervised Medical Image Segmentation with Co-Distribution Alignment
- A(DP)SGD: Asynchronous Decentralized Parallel Stochastic Gradient Descent with Differential Privacy
- Low-cost and high-performance data augmentation for deep-learning-based skin lesion classification
- Making EfficientNet More Efficient: Exploring Batch-Independent Normalization, Group Convolutions and Reduced Resolution Training
- Local Adaptation Improves Accuracy of Deep Learning Model for Automated X-Ray Thoracic Disease Detection : A Thai Study
- S3NAS: Fast NPU-aware Neural Architecture Search Methodology
- For self-supervised learning, Rationality implies generalization, provably
- DeepGalaxy: Deducing the Properties of Galaxy Mergers from Images Using Deep Neural Networks
- Rethinking Recurrent Neural Networks and Other Improvements for Image Classification
- SEM-CLIP: Precise Few-Shot Learning for Nanoscale Defect Detection in Scanning Electron Microscope Image
- Deep learning-based identification of precipitation clouds from all-sky camera data for observatory safety
- Neural Architecture Search for Generative Adversarial Networks: A Comprehensive Review and Critical Analysis
- Term Revealing: Furthering Quantization at Run Time on Quantized DNNs
- A multimodal dataset for understanding the impact of mobile phones on remote online virtual education
- Trusted Media Challenge Dataset and User Study
- A Deep Learning Approach to Teeth Segmentation and Orientation from Panoramic X-rays
- Facing the Hard Problems in FGVC
- Powering One-shot Topological NAS with Stabilized Share-parameter Proxy
- Hazard Detection in Supermarkets using Deep Learning on the Edge
- ESAI: Efficient Split Artificial Intelligence via Early Exiting Using Neural Architecture Search
- Rethinking Training from Scratch for Object Detection
- A Novel lightweight Convolutional Neural Network, ExquisiteNetV2
- K-shot NAS: Learnable Weight-Sharing for NAS with K-shot Supernets
- CAP-GAN: Towards Adversarial Robustness with Cycle-consistent Attentional Purification
- On the Privacy Risks of Cell-Based NAS Architectures
- 1st Place Solution to Google Landmark Retrieval 2020
- Meta-Learning Initializations for Image Segmentation
- The 1st Agriculture-Vision Challenge: Methods and Results
- The 2ST-UNet for Pneumothorax Segmentation in Chest X-Rays using ResNet34 as a Backbone for U-Net
- Improving Accuracy of Binary Neural Networks using Unbalanced Activation Distribution
- FNAS: Uncertainty-Aware Fast Neural Architecture Search
- Learning in Convolutional Neural Networks Accelerated by Transfer Entropy
- EfficientNet Algorithm for Classification of Different Types of Cancer
- Exploring CNN-based models for image's aesthetic score prediction with using ensemble
- DKM: Differentiable K-Means Clustering Layer for Neural Network Compression
- Explainable Autonomous Robots: A Survey and Perspective
- Detecting Transaction-based Tax Evasion Activities on Social Media Platforms Using Multi-modal Deep Neural Networks
- Arithmetic-Intensity-Guided Fault Tolerance for Neural Network Inference on GPUs
- Robust Training of Social Media Image Classification Models for Rapid Disaster Response
- UPANets: Learning from the Universal Pixel Attention Networks
- Edge Bias in Federated Learning and its Solution by Buffered Knowledge Distillation
- Hybrid Composition with IdleBlock: More Efficient Networks for Image Recognition
- Pufferfish: Communication-efficient Models At No Extra Cost
- A Study on Encodings for Neural Architecture Search
- GTNet:Guided Transformer Network for Detecting Human-Object Interactions
- Scalable Verification of Quantized Neural Networks (Technical Report)
- An empirical study of pretrained representations for few-shot classification
- Asymmetric Rejection Loss for Fairer Face Recognition
- Does Data Augmentation Benefit from Split BatchNorms
- Fine-Grained Stochastic Architecture Search
- SAIA: Split Artificial Intelligence Architecture for Mobile Healthcare System
- Approximations in Deep Learning
- A study of CNN capacity applied to Left Venticle Segmentation in Cardiac MRI
- On the Distributional Properties of Adaptive Gradients
- Task-relevant Representation Learning for Networked Robotic Perception
- DeepRA: Predicting Joint Damage From Radiographs Using CNN with Attention
- On Calibration of Mixup Training for Deep Neural Networks
- Towards Latency-aware DNN Optimization with GPU Runtime Analysis and Tail Effect Elimination
- Deep Learning for Rheumatoid Arthritis: Joint Detection and Damage Scoring in X-rays
- ElasticHash: Semantic Image Similarity Search by Deep Hashing with Elasticsearch
- Jointly Optimizing Preprocessing and Inference for DNN-based Visual Analytics
- SparseDNN: Fast Sparse Deep Learning Inference on CPUs
- Exploring the Uncertainty Properties of Neural Networks' Implicit Priors in the Infinite-Width Limit
- Learning an Efficient Network for Large-Scale Hierarchical Object Detection with Data Imbalance: 3rd Place Solution to Open Images Challenge 2019
- Investigating Transfer Learning Capabilities of Vision Transformers and CNNs by Fine-Tuning a Single Trainable Block
- Random Shadows and Highlights: A new data augmentation method for extreme lighting conditions
- Local Concept Embeddings for Analysis of Concept Distributions in Vision DNN Feature Spaces
- WeightNet: Revisiting the Design Space of Weight Networks
- Deep Learning-Based Transfer Learning for Classification of Cassava Disease
- Automatic Analysis System of Calcaneus Radiograph: Rotation-Invariant Landmark Detection for Calcaneal Angle Measurement, Fracture Identification and Fracture Region Segmentation
- Effective Data Fusion with Generalized Vegetation Index: Evidence from Land Cover Segmentation in Agriculture
- Nonlinear Regression with a Convolutional Encoder-Decoder for Remote Monitoring of Surface Electrocardiograms
- ResNetX: a more disordered and deeper network architecture
- Feature Products Yield Efficient Networks
- AIO-P: Expanding Neural Performance Predictors Beyond Image Classification
- Deep-n-Cheap: An Automated Search Framework for Low Complexity Deep Learning
- Analysis of Dimensional Influence of Convolutional Neural Networks for Histopathological Cancer Classification
- ExplainFix: Explainable Spatially Fixed Deep Networks
- Auto-Agent-Distiller: Towards Efficient Deep Reinforcement Learning Agents via Neural Architecture Search
- Self-supervised Model Based on Masked Autoencoders Advance CT Scans Classification
- Swift for TensorFlow: A portable, flexible platform for deep learning
- Recognizing Multiple Ingredients in Food Images Using a Single-Ingredient Classification Model
- Data-Efficient Deep Learning Method for Image Classification Using Data Augmentation, Focal Cosine Loss, and Ensemble
- 3rd Place Solution to "Google Landmark Retrieval 2020"
- Efficient Method for Categorize Animals in the Wild
- HAO: Hardware-aware neural Architecture Optimization for Efficient Inference
- Extending Label Smoothing Regularization with Self-Knowledge Distillation
- Attention-Based Face AntiSpoofing of RGB Images, using a Minimal End-2-End Neural Network
- SuperOCR: A Conversion from Optical Character Recognition to Image Captioning
- PatchNet -- Short-range Template Matching for Efficient Video Processing
- AutoScale: Optimizing Energy Efficiency of End-to-End Edge Inference under Stochastic Variance
- Impact of LiDAR visualisations on semantic segmentation of archaeological objects
- One-class Steel Detector Using Patch GAN Discriminator for Visualising Anomalous Feature Map
- Navigating limitations with precision: A fine-grained ensemble approach to wrist pathology recognition on a limited x-ray dataset
- Auto-NBA: Efficient and Effective Search Over the Joint Space of Networks, Bitwidths, and Accelerators
- A Feasible Framework for Arbitrary-Shaped Scene Text Recognition
- Training Sparse Neural Networks using Compressed Sensing
- Identification of primary angle-closure on AS-OCT images with Convolutional Neural Networks
- A Light-Weighted Convolutional Neural Network for Bitemporal SAR Image Change Detection
- Transfer Learning with Pre-trained Conditional Generative Models
- Prioritized Architecture Sampling with Monto-Carlo Tree Search
- FP-NAS: Fast Probabilistic Neural Architecture Search
- Deep Learning Classification of Lake Zooplankton
- Neural Architecture Search with Random Labels
- RapidRead: Global Deployment of State-of-the-art Radiology AI for a Large Veterinary Teleradiology Practice
- Generative Design of Hardware-aware DNNs
- Balanced Binary Neural Networks with Gated Residual
- G-DARTS-A: Groups of Channel Parallel Sampling with Attention
- FixNorm: Dissecting Weight Decay for Training Deep Neural Networks
- Multi-Feature Semi-Supervised Learning for COVID-19 Diagnosis from Chest X-ray Images
- Weight Equalizing Shift Scaler-Coupled Post-training Quantization
- On the Orthogonality of Knowledge Distillation with Other Techniques: From an Ensemble Perspective
- LiDAR ICPS-net: Indoor Camera Positioning based-on Generative Adversarial Network for RGB to Point-Cloud Translation
- CAMRI Loss: Improving Recall of a Specific Class without Sacrificing Accuracy
- A Survey of Large-Scale Deep Learning Serving System Optimization: Challenges and Opportunities
- Multi-level Feature Fusion-based CNN for Local Climate Zone Classification from Sentinel-2 Images: Benchmark Results on the So2Sat LCZ42 Dataset
- Improving task-specific representation via 1M unlabelled images without any extra knowledge
- Learning and Exploiting Interclass Visual Correlations for Medical Image Classification
- Self-Reorganizing and Rejuvenating CNNs for Increasing Model Capacity Utilization
- Greedy Network Enlarging
- GREEN: a Graph REsidual rE-ranking Network for Grading Diabetic Retinopathy
- Scale Calibrated Training: Improving Generalization of Deep Networks via Scale-Specific Normalization
- WixUp: A General Data Augmentation Framework for Wireless Perception in Tracking of Humans
- Efficient Pre-trained Features and Recurrent Pseudo-Labeling in Unsupervised Domain Adaptation
- Regression on Deep Visual Features using Artificial Neural Networks (ANNs) to Predict Hydraulic Blockage at Culverts
- CRL: Class Representative Learning for Image Classification
- Learning Purified Feature Representations from Task-irrelevant Labels
- Control of Computer Pointer Using Hand Gesture Recognition in Motion Pictures
- A simulator-based autoencoder for focal plane wavefront sensing
- Bridging the Reality Gap for Pose Estimation Networks using Sensor-Based Domain Randomization
- Combining Semantic Guidance and Deep Reinforcement Learning For Generating Human Level Paintings
- CLAWS: Contrastive Learning with hard Attention and Weak Supervision
- Deep regression for uncertainty-aware and interpretable analysis of large-scale body MRI
- Efficient Model Performance Estimation via Feature Histories
- Exceeding the Limits of Visual-Linguistic Multi-Task Learning
- How Convolutional Neural Network Architecture Biases Learned Opponency and Colour Tuning
- AIPerf: Automated machine learning as an AI-HPC benchmark
- Enabling Data Diversity: Efficient Automatic Augmentation via Regularized Adversarial Training
- Multiple Myeloma Cancer Cell Instance Segmentation
- Neural Inheritance Relation Guided One-Shot Layer Assignment Search
- Deep Learning for Prostate Pathology
- The Heterogeneity Hypothesis: Finding Layer-Wise Differentiated Network Architectures
- AutoShrink: A Topology-aware NAS for Discovering Efficient Neural Architecture
- Data Augmentation via Structured Adversarial Perturbations
- AGNet: Weighing Black Holes with Machine Learning
- Pretraining Neural Architecture Search Controllers with Locality-based Self-Supervised Learning
- Transfer Learning using Neural Ordinary Differential Equations
- Covid-19 classification with deep neural network and belief functions
- Image-based Detection of Surface Defects in Concrete during Construction
- A Comparative Study of U-Net Topologies for Background Removal in Histopathology Images
- Neighbourhood Distillation: On the benefits of non end-to-end distillation
- Scaling Laws for the Few-Shot Adaptation of Pre-trained Image Classifiers
- Effective Model Compression via Stage-wise Pruning
- Survey on Deep Learning-based Kuzushiji Recognition
- Guided Evolution for Neural Architecture Search
- Ink Marker Segmentation in Histopathology Images Using Deep Learning
- Depth-Enhanced Feature Pyramid Network for Occlusion-Aware Verification of Buildings from Oblique Images
- Towards All-around Knowledge Transferring: Learning From Task-irrelevant Labels
- Exploring Adversarial Fake Images on Face Manifold
- SmartDeal: Re-Modeling Deep Network Weights for Efficient Inference and Training
- TAL EmotioNet Challenge 2020 Rethinking the Model Chosen Problem in Multi-Task Learning
- AET-EFN: A Versatile Design for Static and Dynamic Event-Based Vision
- Volumization as a Natural Generalization of Weight Decay
- Conditional Neural Architecture Search
- Bosch Deep Learning Hardware Benchmark
- Self-Supervised Neural Architecture Search for Multimodal Deep Neural Networks
- Robust Vision Challenge 2020 -- 1st Place Report for Panoptic Segmentation
- Pediatric Bone Age Prediction Using Deep Learning
- Architectural Resilience to Foreground-and-Background Adversarial Noise
- Third ArchEdge Workshop: Exploring the Design Space of Efficient Deep Neural Networks
- Mastering Large Scale Multi-label Image Recognition with high efficiency overCamera trap images
- Exploring Knowledge Distillation of a Deep Neural Network for Multi-Script identification
- Accelerating Neural Network Inference by Overflow Aware Quantization
- Width Transfer: On the (In)variance of Width Optimization
- Cyclic orthogonal convolutions for long-range integration of features
- Normalized Label Distribution: Towards Learning Calibrated, Adaptable and Efficient Activation Maps
- Spatial light interference microscopy (SLIM): principle and applications to biomedicine
- A Neuro-Inspired Autoencoding Defense Against Adversarial Perturbations
- Itsy Bitsy SpiderNet: Fully Connected Residual Network for Fraud Detection
- Tips and Tricks for Webly-Supervised Fine-Grained Recognition: Learning from the WebFG 2020 Challenge
- Finding Closure: A Closer Look at the Gestalt Law of Closure in Convolutional Neural Networks
- Deep Tiny Network for Recognition-Oriented Face Image Quality Assessment
- Transfer Learning Between Different Architectures Via Weights Injection
- Unchain the Search Space with Hierarchical Differentiable Architecture Search
- Sequential Feature Filtering Classifier
- Improve Global Glomerulosclerosis Classification with Imbalanced Data using CircleMix Augmentation
- ODE guided Neural Data Augmentation Techniques for Time Series Data and its Benefits on Robustness
- Benchmarking down-scaled (not so large) pre-trained language models
- A distillation based approach for the diagnosis of diseases
- AsymmNet: Towards ultralight convolution neural networks using asymmetrical bottlenecks
- DeepActsNet: Spatial and Motion features from Face, Hands, and Body Combined with Convolutional and Graph Networks for Improved Action Recognition
- Deep Neural Networks with Short Circuits for Improved Gradient Learning
- Understanding in Artificial Intelligence
- Exploiting Features with Split-and-Share Module
- Multipath CNN with alpha matte inference for knee tissue segmentation from MRI
- Distractor-Aware Neuron Intrinsic Learning for Generic 2D Medical Image Classifications
- AutoFL: Enabling Heterogeneity-Aware Energy Efficient Federated Learning
- Now that I can see, I can improve: Enabling data-driven finetuning of CNNs on the edge
- Mutually-aware Sub-Graphs Differentiable Architecture Search
- Efficient Data-specific Model Search for Collaborative Filtering