A Quantitative Insight Into the Role of Skip Connections in Deep Neural Networks of Low Complexity: A Case Study Directed at Fluid Flow Modeling

Choubineh, Abouzar;Chen, Jie;Coenen, Frans;Ma, Fei

contributor author	Choubineh, Abouzar;Chen, Jie;Coenen, Frans;Ma, Fei
date accessioned	2022-12-27T23:12:58Z
date available	2022-12-27T23:12:58Z
date copyright	7/18/2022 12:00:00 AM
date issued	2022
identifier issn	1530-9827
identifier other	jcise_23_1_014502.pdf
identifier uri	http://yetl.yabesh.ir/yetl1/handle/yetl/4288130
description abstract	Deep feed-forward networks, with high complexity, backpropagate the gradient of the loss function from final layers to earlier layers. As a consequence, the “gradient” may descend rapidly toward zero. This is known as the vanishing gradient phenomenon that prevents earlier layers from benefiting from further training. One of the most efficient techniques to solve this problem is using skip connection (shortcut) schemes that enable the gradient to be directly backpropagated to earlier layers. This paper investigates whether skip connections significantly affect the performance of deep neural networks of low complexity or whether their inclusion has little or no effect. The analysis was conducted using four Convolutional Neural Networks (CNNs) to predict four different multiscale basis functions for the mixed Generalized Multiscale Finite Element Method (GMsFEM). These models were applied to 249,375 samples. Three skip connection schemes were added to the base structure: Scheme 1 from the first convolutional block to the last, Scheme 2 from the middle to the last block, and Scheme 3 from the middle to the last and the second-to-last blocks. The results demonstrate that the third scheme is most effective, as it increases the coefficient of determination (R2) value by 0.0224–0.044 and decreases the Mean Squared Error (MSE) value by 0.0027–0.0058 compared to the base structure. Hence, it is concluded that enriching the last convolutional blocks with the information hidden in neighboring blocks is more effective than enriching using earlier convolutional blocks near the input layer.
publisher	The American Society of Mechanical Engineers (ASME)
title	A Quantitative Insight Into the Role of Skip Connections in Deep Neural Networks of Low Complexity: A Case Study Directed at Fluid Flow Modeling
type	Journal Paper
journal volume	23
journal issue	1
journal title	Journal of Computing and Information Science in Engineering
identifier doi	10.1115/1.4054868
journal fristpage	14502
journal lastpage	14502_9
page	9
tree	Journal of Computing and Information Science in Engineering:;2022:;volume( 023 ):;issue: 001
contenttype	Fulltext

Files in this item

Name:: jcise_23_1_014502.pdf
Size:: 1.263Mb
Format:: PDF

View/Open

This item appears in the following Collection(s)

Journal of Computing and Information Science in Engineering

Show simple item record

YaBeSH Engineering and Technology Library

Archive

A Quantitative Insight Into the Role of Skip Connections in Deep Neural Networks of Low Complexity: A Case Study Directed at Fluid Flow Modeling

Files in this item

This item appears in the following Collection(s)