Title
Parallel processing for stepwise generalisation method on multi-core PC cluster
Abstract
An approximate query, which is an approximate pattern matching in sequence databases, is one of the most important techniques for many different areas, such as computational biology, text mining, web intelligence and pattern recognition; it returns many similar sub-sequences. In this paper, we refer to a set of such similar sub-sequences as a mismatch cluster. To support users who execute an approximate query on a sequence database to find the regularities of approximate patterns that similar to the query pattern, we have developed the stepwise generalisation method that extracts a reduced expression, called a minimum generalised set, from a mismatch cluster. This paper proposes a novel parallelisation model with a hierarchical task pool for the parallel processing of the stepwise generalisation method on a multi-core PC cluster. To manage tasks efficiently on multi-core CPUs, the proposed model uses the hierarchical task pool and an efficient hierarchical dynamic load balancing technique. We evaluate the proposed method using real protein sequences on an actual multi-core PC cluster. Experimental results confirm that the proposed method performs well on multi-core CPUs and on a multi-core PC cluster.
Year
DOI
Venue
2012
10.1504/IJKWI.2012.050282
I. J. Knowledge and Web Intelligence
Keywords
Field
DocType
approximate query,approximate pattern,mismatch cluster,similar sub-sequences,actual multi-core pc cluster,parallel processing,multi-core cpus,multi-core pc cluster,hierarchical task pool,stepwise generalisation method,text mining,multi core
Data mining,Web intelligence,Sequence database,Generalization,Computer science,Parallel processing,Dynamic load balancing,Multi-core processor,Pattern matching
Journal
Volume
Issue
Citations 
3
2
0
PageRank 
References 
Authors
0.34
20
3
Name
Order
Citations
PageRank
Shinpei Yagi100.34
Keiichi Tamura23713.86
H. Kitakami39449.68