K-means based approach for categorical and non categorical data sets
2014
0 views
0 downloads
Advisor: Doç. Dr. Gökhan Silahtaroğlu
Abstract (EN)
Nowadays, corporations and enterprises are used data mining to increase sales and profits through reaching data. Therefore new algorithms and methods are developed in data minig. The machine learning is indispensable component of data mining. In machine learning, there are a lot of algorithms for classification, clustering etc. One of the well known algorithm is K-Means algorithm in machine learning. In K-Means algorithm non categorical data sets are clustering. However in real world categorical and non categorical data sets are nested. The aim of thesis is to develop K-Means algorithm which does clusters categorical and non categorical data sets together. To do this, Jaccard similarity measure is embeded inside K-Means algorithm instead of Euclid for categorical part of data sets then two algorithms are combined each other clustering categorical and non categorical data sets. Key Words: K-Means, Jaccard Similarity Measure, categorical data, non categorical data
Author
Dr. Mustafa Demirkan
How to Cite
Mustafa Demirkan (Master Thesis). K-means based approach for categorical and non categorical data sets, 2014, İstanbul Beykent Üniversity.
Keywords
License
Tüm Hakları Saklıdır
This work is shared under the specified license terms.
More theses from İstanbul Beykent Üniversity
- Foreclosure of hypothec(2023)
- Evaluation of pre-consumer waste in the Turkish ready-to-wear sector through sustainable design methods(2025)
- Analysis of space identity through corprate identity in historical buildings(2018)
- Use of facial recognition systems in hospital automation system(2018)
- Solution of a production assembly line problem with heuristic methods(2018)
- Inventory management and spare part's stock implementation in a automobile company(2018)
