Data Warehouse

In computing, a data warehouse or enterprise data warehouse (DW, DWH, or EDW) is a database used for reporting and data analysis. It is a central repository of data which is created by integrating data from multiple disparate sources. Data warehouses store current as well as historical data and are used for creating trending reports for senior management reporting such as annual and quarterly comparisons.

The data stored in the warehouse are uploaded from the operational systems (such as marketing, sales etc., shown in the figure to the right). The data may pass through an operational data store for additional operations before they are used in the DW for reporting.

The typical ETL-based data warehouse uses staging, integration, and access layers to house its key functions. The staging layer or staging database stores raw data extracted from each of the disparate source data systems. The integration layer integrates the disparate data sets by transforming the data from the staging layer often storing this transformed data in an operational data store (ODS) database. The integrated data are then moved to yet another database, often called the data warehouse database, where the data is arranged into hierarchical groups often called dimensions and into facts and aggregate facts. The combination of facts and dimensions is sometimes called a star schema. The access layer helps users retrieve data.

A data warehouse constructed from an integrated data source systems does not require ETL, staging databases, or operational data store databases. The integrated data source systems may be considered to be a part of a distributed operational data store layer. Data federation methods or data virtualization methods may be used to access the distributed integrated source data systems to consolidate and aggregate data directly into the data warehouse database tables. Unlike the ETL-based data warehouse, the integrated source data systems and the data warehouse are all integrated since there is no transformation of dimensional or reference data. This integrated data warehouse architecture supports the drill down from the aggregate data of the data warehouse to the transactional data of the integrated source data systems.

Data warehouses can be subdivided into data marts. Data marts store subsets of data from a warehouse.

This definition of the data warehouse focuses on data storage. The main source of the data is cleaned, transformed, cataloged and made available for use by managers and other business professionals for data mining, online analytical processing, market research and decision support (Marakas & O'Brien 2009). However, the means to retrieve and analyze data, to extract, transform and load data, and to manage the data dictionary are also considered essential components of a data warehousing system. Many references to data warehousing use this broader context. Thus, an expanded definition for data warehousing includes business intelligence tools, tools to extract, transform and load data into the repository, and tools to manage and retrieve metadata.

Read more about Data Warehouse:  Benefits of A Data Warehouse, A Generic Data Warehouse Environment, History, Dimensional Vs. Normalized Approach For Storage of Data, Data Warehouses Versus Operational Systems, Evolution in Organization Use, Sample Applications

Other articles related to "data, data warehouse, data warehouses":

Market Intelligence
... Market intelligence includes gathering of data from the company’s external environment, whereas the Business Intelligence process primarily is based on internal recorded events – such as sales, shipments and ... of an Extract, Transform and Load (ETL) processor, a data warehouse and a range of reporting tools ... Information from existing corporate data sources is extracted, transformed and then loaded into a data warehouse ...
Data Warehouse - Sample Applications
... Some of the applications of data warehousing include Agriculture Biological data analysis Call record analysis Churn Prediction for Telecom subscribers ...
Accounting Intelligence Compared To Business Intelligence
... some key ways- Accounting intelligence applications are specifically designed to analyse data from specific ERP systems ... In an accounting intelligence, there is no staggering of data in a data warehouse or OLAP cube ... involves a batch process to extract data from the live database, and store it in a denormalised form ...
Examples of Intelligent Software Agents - Data-mining Agents
... See also Data mining agents This agent uses information technology to find trends and patterns in an abundance of information from many different sources ... A data mining agent operates in a data warehouse discovering information ... A 'data warehouse' brings together information from lots of different sources ...
... performs all the tasks normally done by a data mart or data warehouse, such as extracting, transforming, and formatting data as well as defining metrics, submitting ... Also known as data shadow systems, human data warehouses, or IT shadow systems ... Critics like Stephen Samild argue that the definition stems from a biased view that sees a Data Warehouse as desirable end-result, whereas One might more accurately define data ...

Famous quotes containing the word data:

    To write it, it took three months; to conceive it three minutes; to collect the data in it—all my life.
    F. Scott Fitzgerald (1896–1940)