Relational joins for data on tertiary storage

被引:1
|
作者
Myllymaki, J
Livny, M
机构
关键词
D O I
10.1109/ICDE.1997.581749
中图分类号
TP3 [计算技术、计算机技术];
学科分类号
0812 ;
摘要
Despite the steady decrease in secondary storage prices, the data storage requirements of many organizations cannot be met economically using secondary storage alone. Tertiary storage offers a lower-cost alternative but is viewed as a second-class citizen in many systems. For instance, the typical solution in bringing tertiary-resident data under the control of a DBMS is to use operating system facilities to copy the data to secondary storage, and then to perform query optimization and execution as if the data had been in secondary storage all along. This approach fails to recognize the opportunities for saving execution time and storage space if the data were accessed directly on tertiary devices and in parallel with other I/Os. In this paper we explore how to join two DBMS relations stored on magnetic tapes. Both relations are assumed to be larger than available disk space. We show how Grace Hash Join can be modified to handle a range of tape relation sizes. The modified algorithms access data directly on tapes and exploit parallelism between disk and tape I/Os. We also provide performance results of an experimental implementation of the algorithms.
引用
收藏
页码:159 / 168
页数:10
相关论文
共 50 条
  • [1] Worst Case Optimal Joins on Relational and XML data
    Chen, Yuxing
    [J]. SIGMOD'18: PROCEEDINGS OF THE 2018 INTERNATIONAL CONFERENCE ON MANAGEMENT OF DATA, 2018, : 1833 - 1835
  • [2] A GENERALIZATION OF RELATIONAL JOINS
    KARPOV, AO
    [J]. PROGRAMMING AND COMPUTER SOFTWARE, 1989, 15 (02) : 64 - 70
  • [3] Generalization of relational joins
    Karpov, A.O.
    [J]. Programming and computer software, 1990, 15 (02) : 64 - 70
  • [4] New algorithms for parallelizing relational database joins in the presence of data skew
    [J]. Wolf, Joel L., 1600, IEEE, Piscataway, NJ, United States (06):
  • [5] NEW ALGORITHMS FOR PARALLELIZING RELATIONAL DATABASE JOINS IN THE PRESENCE OF DATA SKEW
    WOLF, JL
    DIAS, DM
    YU, PS
    TUREK, J
    [J]. IEEE TRANSACTIONS ON KNOWLEDGE AND DATA ENGINEERING, 1994, 6 (06) : 990 - 997
  • [6] Tiny disc system joins the data storage fray
    Jensen, GE
    [J]. ELECTRONIC PRODUCTS MAGAZINE, 2001, 43 (09): : 26 - 26
  • [7] Deductive Optimization of Relational Data Storage
    Feser, John
    Madden, Sam
    Tang, Nan
    Solar-Lezama, Armando
    [J]. PROCEEDINGS OF THE ACM ON PROGRAMMING LANGUAGES-PACMPL, 2020, 4
  • [8] Relational Joins on GPUs: A Closer Look
    Yabuta, Makoto
    Anh Nguyen
    Kato, Shinpei
    Edahiro, Masato
    Kawashima, Hideyuki
    [J]. IEEE TRANSACTIONS ON PARALLEL AND DISTRIBUTED SYSTEMS, 2017, 28 (09) : 2663 - 2673
  • [9] Similarity Joins in Relational Database Systems
    Augsten, Nikolaus
    Böhlen, Michael H.
    [J]. Synthesis Lectures on Data Management, 2014, 5 (05): : 1 - 106
  • [10] THE STUDY OF JOINS IN FUZZY RELATIONAL DATABASES
    RAJU, KVSVN
    MAJUMDAR, AK
    [J]. FUZZY SETS AND SYSTEMS, 1987, 21 (01) : 19 - 34