Javascript must be enabled to continue!
Analysis of the K-Means Algorithm for Clustering School Participation Rates in Central Java
View through CrossRef
One indication of the development of educational services in Indonesia is the School Enrollment Rate (SER). Higher the rate of enrolment, the better a location offers access to training. The dataset source was collected from the Central Java Statistical Agency website. The analysis object is the percentage of SERs for ages 7-12 years, 13-15 years, and 16-18 years in the Central Java region during 2017-2019. In the Central Java province, the aim of which is the third largest province after West Java and East Java, was to analyze the level of school participation as mapped. The created research product is a mapping of locations in the District and City areas in the form of clusters. The solution is the clustering algorithm k-means. In this study, there were two groups: high (C1) and low. The clusters were separated into (C2). Cluster-mapping studies results for the years 7-12 were, that in a high cluster, 24 provinces (cluster 0) and 11 provinces (cluster 1) were in a lower cluster, whereas the 13-15-year-old cluster mapping results from 23 provinces (cluster 0) and 12 provinces (cluster 1) and the 16-18-year-old cluster mapping results from 15 provinces. Final centroid value is the basis for the determination of the clusters where the final centroid value for a cluster aged 7-12 years were high (cluster 0) {99.81, 99.87, 99.75} and low (cluster 1) {99.73, 99.43, 99.25}, whereas the final centroid value of a cluster aged 13-15 years was high (cluster 0). For all age categories, the mapping findings reveal a good proportion, that is, over 50% in the top class. In particular, 24 provinces (57%) were in the low cluster of the 16-18-year age group. Research results information can provide a macro-image of the level of SER development in recent years.
Keywords: K-Means, algorithm, clustering
Title: Analysis of the K-Means Algorithm for Clustering School Participation Rates in Central Java
Description:
One indication of the development of educational services in Indonesia is the School Enrollment Rate (SER).
Higher the rate of enrolment, the better a location offers access to training.
The dataset source was collected from the Central Java Statistical Agency website.
The analysis object is the percentage of SERs for ages 7-12 years, 13-15 years, and 16-18 years in the Central Java region during 2017-2019.
In the Central Java province, the aim of which is the third largest province after West Java and East Java, was to analyze the level of school participation as mapped.
The created research product is a mapping of locations in the District and City areas in the form of clusters.
The solution is the clustering algorithm k-means.
In this study, there were two groups: high (C1) and low.
The clusters were separated into (C2).
Cluster-mapping studies results for the years 7-12 were, that in a high cluster, 24 provinces (cluster 0) and 11 provinces (cluster 1) were in a lower cluster, whereas the 13-15-year-old cluster mapping results from 23 provinces (cluster 0) and 12 provinces (cluster 1) and the 16-18-year-old cluster mapping results from 15 provinces.
Final centroid value is the basis for the determination of the clusters where the final centroid value for a cluster aged 7-12 years were high (cluster 0) {99.
81, 99.
87, 99.
75} and low (cluster 1) {99.
73, 99.
43, 99.
25}, whereas the final centroid value of a cluster aged 13-15 years was high (cluster 0).
For all age categories, the mapping findings reveal a good proportion, that is, over 50% in the top class.
In particular, 24 provinces (57%) were in the low cluster of the 16-18-year age group.
Research results information can provide a macro-image of the level of SER development in recent years.
Keywords: K-Means, algorithm, clustering.
Related Results
The Kernel Rough K-Means Algorithm
The Kernel Rough K-Means Algorithm
Background:
Clustering is one of the most important data mining methods. The k-means
(c-means ) and its derivative methods are the hotspot in the field of clustering research in re...
Macroeconomic and Social Precursors of Suicide Rates in the Philippines: A Quantitative Analysis (Preprint)
Macroeconomic and Social Precursors of Suicide Rates in the Philippines: A Quantitative Analysis (Preprint)
BACKGROUND
Suicide is a complex, serious and multifaceted public health issue that poses significant challenges to societies worldwide. In fact, it represen...
Wyniki badań 110 dziewcząt “nie uczących się i nie pracujących”
Wyniki badań 110 dziewcząt “nie uczących się i nie pracujących”
The publication presents the findings of an inquiry conducted among 110 girls aged 15 - 17 who had been directed, on the grounds of being “out of school and out of work”, to two on...
ANALYSIS OF THE APPLICABILITY CRITERION FOR K MEANS CLUSTERING ALGORITHM RUN TEN NUMBER OF TIMES ON THE FIRST 25 NUMBERS OF THE FIBONACCI SERIES
ANALYSIS OF THE APPLICABILITY CRITERION FOR K MEANS CLUSTERING ALGORITHM RUN TEN NUMBER OF TIMES ON THE FIRST 25 NUMBERS OF THE FIBONACCI SERIES
In this research investigation Analysis Of The Applicability Criterion For K Means Clustering Algorithm Run Ten Number Of Times On The First 25 Numbers Of The Fibonacci Series is p...
Parallel density clustering algorithm based on MapReduce and optimized cuckoo algorithm
Parallel density clustering algorithm based on MapReduce and optimized cuckoo algorithm
In the process of parallel density clustering, the boundary points of clusters with different densities are blurred and there is data noise, which affects the clustering performanc...
A Hybrid K-means Method based on Modified Rat Swarm Optimization Algorithm for Data Clustering
A Hybrid K-means Method based on Modified Rat Swarm Optimization Algorithm for Data Clustering
Abstract
The original K-means clustering algorithm is prone to local optima and sensitive to the initial clustering center, which have a great impact on accuracy and stabil...
Evaluating Clustering Algorithms: An Analysis using the EDAS Method
Evaluating Clustering Algorithms: An Analysis using the EDAS Method
Data clustering is frequently utilized in the early stages of analyzing big data. It enables the examination of massive datasets encompassing diverse types of data, with the aim of...
Big Data Clustering Method Based on an Improved PSO-Means Algorithm
Big Data Clustering Method Based on an Improved PSO-Means Algorithm
There are problems in big data clustering processing, such as poor clustering effect of different types of data and long clustering time. Therefore, a big data clustering processin...

