Skip to main navigation Skip to search Skip to main content

Join algorithm based on data grid

  • Dong Hua Yang*
  • , Jian Zhong Li
  • , Wen Ping Zhang
  • *Corresponding author for this work
  • School of Computer Science and Technology, Harbin Institute of Technology
  • Heilongjiang University

Research output: Contribution to journalArticlepeer-review

Abstract

Data grid is a distributed architecture for data management, which could provide the coordinated management mechanisms for data distributed across remote resources and form a single, virtual environment for data access, management and process by integrating many data sets distributed in the network. Database management system acts as an important role on data grid. In all kinds of database operations, join operation is a common-used and complex operation that needs much more time to complete than other operations. The join algorithm of massive data on data grid is studied in this paper. The proposed algorithm uses relation reduction algorithm, row blocking technique, and pipelined parallelism to solve the problem of heterogeneity of network bandwidth between nodes on data grid. The analysis and experimental results show that the performance of the algorithm is good in minimizing the response time by decreasing network transmission cost and increasing parallelism of I/O and CPU.

Original languageEnglish
Pages (from-to)1848-1855
Number of pages8
JournalJisuanji Yanjiu yu Fazhan/Computer Research and Development
Volume41
Issue number10
StatePublished - Oct 2004
Externally publishedYes

Keywords

  • Data grid
  • Join operation
  • Pipelined parallelism
  • Relation reduction algorithm

Fingerprint

Dive into the research topics of 'Join algorithm based on data grid'. Together they form a unique fingerprint.

Cite this