Javascript must be enabled to continue!
A Distributed Algorithm for Fast Mining Frequent Patterns in Limited and Varying Network Bandwidth Environments
View through CrossRef
Data mining is a set of methods used to mine hidden information from data. It mainly includes frequent pattern mining, sequential pattern mining, classification, and clustering. Frequent pattern mining is used to discover the correlation among various sets of items within large databases. The rapid upward trend in data size slows the mining of frequent patterns. Numerous studies have attempted to develop algorithms that operate in distributed computing environments to accelerate the mining process. FLR-mining (Fast, Load balancing and Resource efficient mining algorithm) is one of the fastest methods of mining with efficient consideration of load balancing and resources. FLR-mining can automatically determine the appropriate number of computing nodes. However, FLR-mining and existing methods assume that the network bandwidth is constant. In practical distributed and many-task computing systems, this assumption fails because there are packet collisions caused by many mining tasks that run in a simultaneous manner. Therefore, a method that can consider the varying network bandwidth is necessary. In this study, we propose a method that can rapidly mine frequent patterns under the varying network bandwidth. The proposed method can also determine the appropriate number of computing nodes to efficiently utilize computing resources and achieve load balancing. Through empirical evaluation, the proposed method is shown to deliver excellent performance in terms of execution efficiency and load balancing.
Title: A Distributed Algorithm for Fast Mining Frequent Patterns in Limited and Varying Network Bandwidth Environments
Description:
Data mining is a set of methods used to mine hidden information from data.
It mainly includes frequent pattern mining, sequential pattern mining, classification, and clustering.
Frequent pattern mining is used to discover the correlation among various sets of items within large databases.
The rapid upward trend in data size slows the mining of frequent patterns.
Numerous studies have attempted to develop algorithms that operate in distributed computing environments to accelerate the mining process.
FLR-mining (Fast, Load balancing and Resource efficient mining algorithm) is one of the fastest methods of mining with efficient consideration of load balancing and resources.
FLR-mining can automatically determine the appropriate number of computing nodes.
However, FLR-mining and existing methods assume that the network bandwidth is constant.
In practical distributed and many-task computing systems, this assumption fails because there are packet collisions caused by many mining tasks that run in a simultaneous manner.
Therefore, a method that can consider the varying network bandwidth is necessary.
In this study, we propose a method that can rapidly mine frequent patterns under the varying network bandwidth.
The proposed method can also determine the appropriate number of computing nodes to efficiently utilize computing resources and achieve load balancing.
Through empirical evaluation, the proposed method is shown to deliver excellent performance in terms of execution efficiency and load balancing.
Related Results
Light at the End of the Tunnel: Mining Justice and Health
Light at the End of the Tunnel: Mining Justice and Health
The mining industry provides valuable mined commodities and financial support for communities worldwide. Mining has become safer for workers. Significant injustices, however, are c...
Exact and Approximate Digraph Bandwidth
Exact and Approximate Digraph Bandwidth
Abstract
Note: Please see pdf for full abstract with equations.
In this paper, we introduce a directed variant of the classical BANDWIDTH problem and study it from the view...
Exact and Approximate Digraph Bandwidth
Exact and Approximate Digraph Bandwidth
Abstract
In this paper, we introduce a directed variant of the classical Bandwidthproblem and study it from the view-point of moderately exponential time algorithms, both...
IMPLEMENTASI HOTSPOT DENGAN PENGELOLAAN USER MANAGER DAN BANDWIDTH MENGGUNAKAN MIKROTIK RB941-2ND (STUDI KASUS SMK KESEHATAN BHAKTI KENCANA JATIWANGI)
IMPLEMENTASI HOTSPOT DENGAN PENGELOLAAN USER MANAGER DAN BANDWIDTH MENGGUNAKAN MIKROTIK RB941-2ND (STUDI KASUS SMK KESEHATAN BHAKTI KENCANA JATIWANGI)
SMK Kesehatan Bhakti Kencana Jatiwangi merupakan sekolah kejuruan swasta yang berada di Kecamatan Jatiwangi Kabupaten Majalengka, SMK Kesehatan Bhakti Kencana Jatiwangi sudah memil...
Link-Based Signalized Arterial Progression Optimization with Practical Travel Speed
Link-Based Signalized Arterial Progression Optimization with Practical Travel Speed
Bandwidth is defined as the maximum amount of green time for a designated movement as it passes through an arterial. In most previous studies, bandwidth has been referred to arteri...
Distributed frequent hierarchical pattern mining for robust and efficient large-scale association discovery
Distributed frequent hierarchical pattern mining for robust and efficient large-scale association discovery
Frequent pattern mining is a classic data mining technique, generally applicable to a wide range of application domains, and a mature area of research. The fundamental challenge ar...
Estimasi Kebutuhan Bandwidth Internet di Jurusan Teknik Elektro Politeknik Negeri Lhokseumawe
Estimasi Kebutuhan Bandwidth Internet di Jurusan Teknik Elektro Politeknik Negeri Lhokseumawe
Bandwidth internet merupakan salah satu parameter utama yang menjadi ukuran oleh pengguna dalam mengakses jaringan internet. Bandwidth yang bagus akan membuat pengguna nyaman dalam...
Mining frequent patterns without candidate generation
Mining frequent patterns without candidate generation
Mining frequent patterns in transaction databases, time-series databases, and many other kinds of databases has been studied popularly in data mining research. Most of the previous...

