Deep network in network

Alaeddine, Hmidi; Malek, Jihene

doi:10.1007/s00521-020-05008-0

Cited by 20 publications

(31 citation statements)

References 30 publications

Supporting

Mentioning

Contrasting

Order By: Relevance

“…However, changing the filter type is an important step to develop efficient CNNs. Using a nonlinear and more complex filter, such as an MLP filter, can generate more interesting results than using a simple linear filter [ 1 , 21 ]. Several architectures were based on this principle such as [ 1 , 21 , 42 ].…”

Section: Related Workmentioning

confidence: 99%

“…Using a nonlinear and more complex filter, such as an MLP filter, can generate more interesting results than using a simple linear filter [ 1 , 21 ]. Several architectures were based on this principle such as [ 1 , 21 , 42 ]. NIN [ 21 ] adopts a nonlinear filter: the multilayer perceptron (MLP) with a rectified linear unit (ReLU) used as an activation function.…”

Section: Related Workmentioning

confidence: 99%

“…NIN [ 21 ] adopts a nonlinear filter: the multilayer perceptron (MLP) with a rectified linear unit (ReLU) used as an activation function. In [ 1 ], DNIN model directly modifies NIN [ 21 ] in the sense of convolutional layer. It is represented in a three-layer stacking DMLPconv, which consists of two convolutional layers of size 3 × 3 and an eLU unit, used as an activation function instead of ReLU.…”

Section: Related Workmentioning

confidence: 99%

“…It is represented in a three-layer stacking DMLPconv, which consists of two convolutional layers of size 3 × 3 and an eLU unit, used as an activation function instead of ReLU. By incorporating micronetwork, DNIN [ 1 ] also increases depth. The depth of DNIN [ 1 ] is the same as that of NIN [ 21 ] and shares the same number of convolutional kernels.…”

Section: Related Workmentioning

confidence: 99%

“…They have had great success and reached the state of the art in several benchmarks. In this article, we address the degradation problem by introducing an efficient deep neural network architecture for computer vision, deep residual network in network, which takes its name from the deep network in the network article [ 1 ] in conjunction with the famous “deep residual learning for image recognition” [ 14 ]. The advantages of the architecture are experimentally verified on the CIFAR-10 classification challenges.…”

Section: Introductionmentioning

confidence: 99%

See 4 more Smart Citations

Deep Residual Network in Network

Alaeddine

Malek

2021

Computational Intelligence and Neuroscience

Self Cite

View full text Add to dashboard Cite

Deep network in network (DNIN) model is an efficient instance and an important extension of the convolutional neural network (CNN) consisting of alternating convolutional layers and pooling layers. In this model, a multilayer perceptron (MLP), a nonlinear function, is exploited to replace the linear filter for convolution. Increasing the depth of DNIN can also help improve classification accuracy while its formation becomes more difficult, learning time gets slower, and accuracy becomes saturated and then degrades. This paper presents a new deep residual network in network (DrNIN) model that represents a deeper model of DNIN. This model represents an interesting architecture for on-chip implementations on FPGAs. In fact, it can be applied to a variety of image recognition applications. This model has a homogeneous and multilength architecture with the hyperparameter “L” (“L” defines the model length). In this paper, we will apply the residual learning framework to DNIN and we will explicitly reformulate convolutional layers as residual learning functions to solve the vanishing gradient problem and facilitate and speed up the learning process. We will provide a comprehensive study showing that DrNIN models can gain accuracy from a significantly increased depth. On the CIFAR-10 dataset, we evaluate the proposed models with a depth of up to L = 5 DrMLPconv layers, 1.66x deeper than DNIN. The experimental results demonstrate the efficiency of the proposed method and its role in providing the model with a greater capacity to represent features and thus leading to better recognition performance.

show abstract

Section: Related Workmentioning

confidence: 99%

Section: Related Workmentioning

confidence: 99%

Section: Related Workmentioning

confidence: 99%

Section: Related Workmentioning

confidence: 99%

Section: Introductionmentioning

confidence: 99%

See 3 more Smart Citations