EfficientNet

Introduced by Tan et al. in EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks

EfficientNet is a convolutional neural network architecture and scaling method that uniformly scales all dimensions of depth/width/resolution using a compound coefficient. Unlike conventional practice that arbitrary scales these factors, the EfficientNet scaling method uniformly scales network width, depth, and resolution with a set of fixed scaling coefficients. For example, if we want to use $2^N$ times more computational resources, then we can simply increase the network depth by $\alpha ^ N$, width by $\beta ^ N$, and image size by $\gamma ^ N$, where $\alpha, \beta, \gamma$ are constant coefficients determined by a small grid search on the original small model. EfficientNet uses a compound coefficient $\phi$ to uniformly scales network width, depth, and resolution in a principled way.

The compound scaling method is justified by the intuition that if the input image is bigger, then the network needs more layers to increase the receptive field and more channels to capture more fine-grained patterns on the bigger image.

The base EfficientNet-B0 network is based on the inverted bottleneck residual blocks of MobileNetV2, in addition to squeeze-and-excitation blocks.

EfficientNets also transfer well and achieve state-of-the-art accuracy on CIFAR-100 (91.7%), Flowers (98.8%), and 3 other transfer learning datasets, with an order of magnitude fewer parameters.

Source: EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks

Read Paper See Code

Papers

Paper	Code	Results	Date	Stars

Tasks

Task	Papers	Share
Image Classification	38	11.59%
General Classification	20	6.10%
Semantic Segmentation	17	5.18%
Classification	16	4.88%
Object Detection	16	4.88%
Decoder	6	1.83%
Multi-Task Learning	6	1.83%
Instance Segmentation	6	1.83%
COVID-19 Diagnosis	5	1.52%

Usage Over Time

This feature is experimental; we are continuously improving our matching algorithm.

Components

Component	Type	Add Remove
1x1 Convolution	Convolutions
Average Pooling	Pooling Operations
Convolution	Convolutions
Dense Connections	Feedforward Networks
Dropout	Regularization
Inverted Residual Block	Skip Connection Blocks
RMSProp	Stochastic Optimization
Squeeze-and-Excitation Block	Image Model Blocks
Swish	Activation Functions

Categories

Add Remove

Image Models

Convolutional Neural Networks