Integrating data warehouses with web data:: A survey

被引:58
|
作者
Manuel Perez, Juan [1 ]
Berlanga, Rafael [1 ]
Jose Aramburu, Maria [2 ]
Pedersen, Torben Bach [3 ]
机构
[1] Univ Jaume 1, Dept Lenguajes & Sistemas Informat, E-12071 Castellon de La Plana, Spain
[2] Univ Jaume 1, Dept Ingn & Ciencia Computadores, E-12071 Castellon de La Plana, Spain
[3] Aalborg Univ, Dept Comp Sci, DK-9220 Aalborg O, Denmark
关键词
data warehouse repository; XML/XSL/RDF;
D O I
10.1109/TKDE.2007.190746
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
This paper surveys the most relevant research on combining Data Warehouse (DW) and Web data. It studies the XML technologies that are currently being used to integrate, store, query, and retrieve Web data and their application to DWs. The paper reviews different DW distributed architectures and the use of XML languages as an integration tool in these systems. It also introduces the problem of dealing with semistructured data in a DW. It studies Web data repositories, the design of multidimensional databases for XML data sources, and the XML extensions of OnLine Analytical Processing techniques. The paper addresses the application of information retrieval technology in a DW to exploit text-rich document collections. The authors hope that the paper will help to discover the main limitations and opportunities that offer the combination of the DW and the Web fields, as well as to identify open research lines.
引用
收藏
页码:940 / 955
页数:16
相关论文
共 50 条
  • [1] Dynamic approach for integrating web data warehouses
    Le, D. Xuan
    Rahayu, J. Wenny
    Pardede, Eric
    [J]. COMPUTATIONAL SCIENCE AND ITS APPLICATIONS - ICCSA 2006, PT 4, 2006, 3983 : 207 - 216
  • [2] Building data warehouses with semantic web data
    Nebot, Victoria
    Berlanga, Rafael
    [J]. DECISION SUPPORT SYSTEMS, 2012, 52 (04) : 853 - 868
  • [3] Integrating Star and Snowflake Schemas in Data Warehouses
    Garani, Georgia
    Helmer, Sven
    [J]. INTERNATIONAL JOURNAL OF DATA WAREHOUSING AND MINING, 2012, 8 (04) : 22 - 40
  • [4] Challenges and Conflicts Integrating Heterogeneous Data Warehouses
    Preis, Marcus
    Seitz, Juergen
    [J]. TENTH WUHAN INTERNATIONAL CONFERENCE ON E-BUSINESS, VOLS I AND II, 2011, : 577 - 585
  • [6] A Survey of Managing the Evolution of Data Warehouses
    Wrembel, Robert
    [J]. INTERNATIONAL JOURNAL OF DATA WAREHOUSING AND MINING, 2009, 5 (02) : 24 - 56
  • [7] A foundation for spatial data warehouses on the Semantic Web
    Gur, Nurefsan
    Pedersen, Torben Bach
    Zimanyi, Esteban
    Hose, Katja
    [J]. SEMANTIC WEB, 2018, 9 (05) : 557 - 587
  • [8] Integrating heterogeneous data warehouses using XML technologies
    Tseng, FSC
    Chen, CW
    [J]. JOURNAL OF INFORMATION SCIENCE, 2005, 31 (03) : 209 - 229
  • [9] Design and development of a tool for integrating heterogeneous data warehouses
    Torlone, R
    Panella, I
    [J]. DATA WAREHOUSING AND KNOWLEDGE DISCOVERY, PROCEEDINGS, 2005, 3589 : 105 - 114
  • [10] Conceptual Multidimensional Modeling for Data Warehouses: A Survey
    Gosain, Anjana
    Singh, Jaspreeti
    [J]. PROCEEDINGS OF THE 3RD INTERNATIONAL CONFERENCE ON FRONTIERS OF INTELLIGENT COMPUTING: THEORY AND APPLICATIONS (FICTA) 2014, VOL 1, 2015, 327 : 305 - 316