Abstract
Query processing in data grids is a difficult issue due to the heterogeneous, unpredictable and volatile behaviors of the grid resources. Applying join operations on remote relations in data grids is a unique and interesting problem. However, to the best of our knowledge, little is done to date on multi-join query processing in data grids. An approach for processing multi-join queries is proposed in this paper. Firstly, a relation-reduction algorithm for reducing the sizes of operand relations is presented in order to minimize data transmission cost among grid nodes. Then, a method for scheduling computer nodes in data grids is devised to parallel process multi-join queries. Thirdly, an innovative method is developed to efficiently execute join operations in a pipeline fashion. Finally, a complete algorithm for processing multi-join queries is given. Analytical and experimental results show the effectiveness and efficiency of the proposed approach.
| Original language | English |
|---|---|
| Pages (from-to) | 3574-3591 |
| Number of pages | 18 |
| Journal | Information Sciences |
| Volume | 177 |
| Issue number | 17 |
| DOIs | |
| State | Published - 1 Sep 2007 |
| Externally published | Yes |
Keywords
- Data grids
- Minimum-maximum-edge matching
- Multi-join query processing
- Relation-reduction algorithm
Fingerprint
Dive into the research topics of 'Distributed multi-join query processing in data grids'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver