An experimental study on hyper parameters for training deep convolutional networks

dc.authorid0000-0002-1351-7565en_US
dc.contributor.authorTemiz, Hakan
dc.date.accessioned2021-01-05T09:31:29Z
dc.date.available2021-01-05T09:31:29Z
dc.date.issued2020
dc.departmentAÇÜ, Borçka Acarlar Meslek Yüksekokuluen_US
dc.description4th International Symposium on Multidisciplinary Studies and Innovative Technologies, ISMSIT 2020; Istanbul; Turkey; 22 October 2020 through 24 October 2020; Category numberCFP20Q88-ART; Code 165025en_US
dc.description.abstractWhen training deep networks, it is crucial to obtain a network that offers optimum performance by trying different values of many hyper parameters and combinations of these values. Theoretically, the optimal set of values, which ensure the maximum performance of the network, can be found by giving these parameters numerous different values. However, it is not feasible to try all combinations of values. The a priori information regarding the contribution levels of hyper parameters and their values to the performance of the network will narrow the search space and enable researchers to easily and quickly obtain the network with optimum performance. In this study, a priori information is investigated that will guide in searching for most important hyper parameters and their ideal values that ensure optimum performance of a typical convolutional neural network in single image super resolution. For this purpose, the importance levels of the 5 most commonly used hyper parameters in training, and their optimum values were investigated. By giving two different values that are widely used or known to give good results from previous works in the literature for each hyper parameter, in total, 32 different training procedure were performed. The results showed that the learning rate has the most important effect on the performance of the network, then normalization, and then the size of the input image given to the model during training. It has also been found that the batch number and step count parameter values do not make a significant change in the performance of the network. The results obtained from this study could help researchers in determining the training parameters and their values in order to efficiently and rapidly obtain optimum network performance
dc.identifier.citationTemiz, H. (2020, October). An Experimental Study on Hyper Parameters for Training Deep Convolutional Networks. In 2020 4th International Symposium on Multidisciplinary Studies and Innovative Technologies (ISMSIT) (pp. 1-8). IEEE.en_US
dc.identifier.doi10.1109/ISMSIT50672.2020.9254621
dc.identifier.scopusqualityN/A
dc.identifier.urihttps://hdl.handle.net/11494/2517
dc.indekslendigikaynakScopus
dc.institutionauthorTemiz, Hakan
dc.language.isoenen_US
dc.publisherInstitute of Electrical and Electronics Engineers Inc.en_US
dc.relation.ispartof4th International Symposium on Multidisciplinary Studies and Innovative Technologies
dc.relation.publicationcategoryKonferans Öğesi - Uluslararası - Kurum Öğretim Elemanıen_US
dc.rightsinfo:eu-repo/semantics/closedAccessen_US
dc.subjectDeep learningen_US
dc.subjectConvolutional neural networken_US
dc.subjectHyper parameteren_US
dc.subjectTrainingen_US
dc.titleAn experimental study on hyper parameters for training deep convolutional networksen_US
dc.typeConference Object

Dosyalar

Orijinal paket
Listeleniyor 1 - 1 / 1
[ X ]
İsim:
hakan_temiz_2020.pdf
Boyut:
636.7 KB
Biçim:
Adobe Portable Document Format
Açıklama:
Lisans paketi
Listeleniyor 1 - 1 / 1
[ X ]
İsim:
license.txt
Boyut:
1.44 KB
Biçim:
Item-specific license agreed upon to submission
Açıklama: