Derin sinir ağlarında oto-kodlayıcının düzenlenmesi için terkinim ve seyreklik kullanımı
2015
0 views
0 downloads
Advisor: Prof. Dr. Ömer Morgül
Abstract (EN)
Deep learning has emerged as an effective pre-training technique for neural networks with many hidden layers. To overcome the over-fi tting issue, usually large capacity models are used. In this thesis, two methodologies which are frequently utilized in deep neural network literature have been considered. Firstly, for pre-training the performance of sparse autoencoder has been improved by adding p-norm of the sparse penalty term to an over-complete case. This effi ciently induces sparsity to the hidden layers of a deep network to overcome over- fitting issues. At the end of the training, features constructed for each layer end up with a variety of useful information to initialize a deep network. The accuracy obtained is comparable to the conventional sparse autoencoder technique. Secondly, the large capacity networks suff er from complex co-adaptations between the hidden layers by combining the predictions of each unit in the previous layer to generate the features of the next layer. This results to certain redundant features. So, the idea we propose is to induce a threshold level on the hidden activations to allow only the highest active units to participate in the reconstruction of the features and suppressing the e ffect of less active units in the optimization. This is implemented by dropping out k-lowest hidden units while retaining the rest. Our simulations confi rm the hypothesis that the k-lowest dropouts help the optimization in both the pre-training and fi ne-tuning phases giving rise to the internal distributed representations for better generalization. Moreover, this model gives quick convergence than the conventional dropout method. In classi fication task on MNIST dataset, the proposed idea gives the comparable results with the previous regularization techniques such as denoising autoencoders, use of rectifi er linear units combined with standard regularizations. The deep networks constructed from the combination of our models achieve favorably the similar state of the art results obtained by dropout idea with less time complexity making them well suited to large problem sizes.
Author
Dr. Muhaddisa Barat Ali
Institution
How to Cite
Muhaddisa Barat Ali (Master Thesis). Derin sinir ağlarında oto-kodlayıcının düzenlenmesi için terkinim ve seyreklik kullanımı, 2015, Bilkent University.
Keywords
License
Tüm Hakları Saklıdır
This work is shared under the specified license terms.
More theses from Bilkent University
- Geç Antik Çağ'da Aşağı Tuna: Histria örneği(2023)
- Petrol fiyatları ve getiri eğrisi(2024)
- Sözle yönlendirme üzerine makaleler(2014)
- İletişim ağları ve sağlık uygulamaları için çok kollu haydut algoritmaları(2022)
- Türk Anayasa Mahkemesinin içtihatları ışığında karşılaştırmalı anayasal mutluluk(2023)
- Doğrusal karbon zincirlerinin yoğunluk fonksiyoneli teorisi ile incelenmesi(2023)
