Digitala Vetenskapliga Arkivet

Change search
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf
Efficient OLAP query processing in distributed data warehouses
SMHI.
Show others and affiliations
2003 (English)In: Information Systems, ISSN 0306-4379, E-ISSN 1873-6076, Vol. 28, no 1-2, p. 111-135Article in journal (Refereed) Published
Abstract [en]

The success of Internet applications has led to an explosive growth in the demand for bandwidth from. Internet Service Providers. Managing an Internet protocol network requires collecting and analyzing network data, such as flow-level traffic statistics. Such analyses can typically be expressed as OLAP queries, e.g., correlated aggregate queries and data cubes. Current day OLAP tools for this task assume the availability of the data in a centralized data warehouse. However, the inherently distributed nature of data collection and the huge amount of data extracted at each collection point make it impractical to gather all data at a centralized site. One solution is to maintain a distributed data warehouse, consisting of local data warehouses at-each collection point and a coordinator site, with most of the processing being performed at the local sites. In this paper, we consider the problem of efficient evaluation of OLAP queries over a distributed data warehouse. We have developed the Skalla system for this task. Skalla translates OLAP queries, specified as certain algebraic expressions, into distributed evaluation plans which are shipped to individual sites. A salient property of our approach is that only partial results are shipped - never parts of the detail data. We propose a variety of optimizations to minimize both the synchronization traffic and the local processing done at each site. We finally present an experimental study based on TPC-R data. Our results demonstrate the scalability of our techniques and quantify the performance benefits of the optimization techniques that have gone into the Skalla system. (C) 2002 Elsevier Science Ltd. All rights reserved.

Place, publisher, year, edition, pages
2003. Vol. 28, no 1-2, p. 111-135
National Category
Meteorology and Atmospheric Sciences
Research subject
Remote sensing
Identifiers
URN: urn:nbn:se:smhi:diva-1354DOI: 10.1016/S0306-4379(02)00051-0ISI: 000180548200006OAI: oai:DiVA.org:smhi-1354DiVA, id: diva2:845487
Available from: 2015-08-12 Created: 2015-07-29 Last updated: 2017-12-04Bibliographically approved

Open Access in DiVA

No full text in DiVA

Other links

Publisher's full text
By organisation
SMHI
In the same journal
Information Systems
Meteorology and Atmospheric Sciences

Search outside of DiVA

GoogleGoogle Scholar

doi
urn-nbn

Altmetric score

doi
urn-nbn
Total: 446 hits
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf