Date archive for October, 2008
-
Dimensional Tables and Fact Tables
One of the secrets to putting together a good set of data marts is the concept of dimensions. There are two key steps being able to analyse your data, and to build a working data mart model.  Build a set of clean, consistent dimension tables that store reference information about your key dimensions like Product, […]
-
Data modelling Hierarchies- how to make a dimension
One of the most useful data model structures in a data mart is a Hierarchy (also called a Tree structure). Tree structures let us take a large number of things and organise them in a way that makes sense. More importantly, a tree structure lets us “drill down†into information.  Hierarchy Rules In a simple tree […]
-
Duplicate Data and removing duplicate records
Duplicate records, doubles, redundant data, duplicate rows; it doesn’t matter what you call them, they are one of the biggest problems in any data analyst’s life. There are lots of different types of data quality problems, but in this post I’ll focus on Duplicates. I’ll share some hints on how to find duplicate records and remove duplicate records, […]
-
Business Intelligence Workspaces and in memory self serve analysis
In the classical Business Intelligence architecture, users sit at their computers, requesting reports and analysis, and huge central servers churn through the numbers. The dual-core machine on the desk with 2Gb of RAM is asked to do almost nothing. As machines get faster and faster, new tools that use memory are going to create functionality and speed that just hasn’t been […]

