Using Kettel for Data Quality & Integration

** You have to take screenshot for every step in the program (Kettel) and send it to me in a separate file please.

** the main instruction for this order **

It is investigation of data integration issues, you have been asked to perform an integration of sample excerpts from the membership details databases of both organisations. The membership databases excerpts only provide a small sample of the total datasets, and are also limited to only a small number of attributes for each customer record. So, you should perform:

1. An initial quality assessment of the datasets provided by both gyms.
2. Design and perform an integration of the datasets
3. Evaluate the integration process through an investigation of resultant dataset quality.

** I will upload 4 cv files to used it.

the integration project should identify:
1. Such overlapping customers, 2. As well as
a. eliminating replication of information within individual gym datasets, and b. providing a master dataset of contact details for the new merged gym.
*
*
*
More specifically, in order to complete this project you will need to:
 Investigate the properties of the three datasets with a Data Quality investigation. You may conduct this investigation with the DataCleaner Kettle plugin or any other kettle step that you find suitable (for example fuzzy match).
 On the basis of this project specification and your initial investigation, you should design and implement a Data Integration project to:
o Eliminate replicated records found in either Globo Gym or Average Joe’s datasets.
o Identify any customers who are potentially shared by both Average Joe’s and Globo Gym
o Create a new master customer list with standardised information on all customers. Standardised here refers to the use of consistent customer information representation on address, date of birth, and so forth.

Order a unique copy of this paper
(550 words)

Approximate price: $22

Basic features
  • Free title page and bibliography
  • Unlimited revisions
  • Plagiarism-free guarantee
  • Money-back guarantee
  • 24/7 support
On-demand options
  • Writer’s samples
  • Part-by-part delivery
  • Overnight delivery
  • Copies of used sources
  • Expert Proofreading
Paper format
  • 275 words per page
  • 12 pt Arial/Times New Roman
  • Double line spacing
  • Any citation style (APA, MLA, Chicago/Turabian, Harvard)

Our guarantees

Delivering a high-quality product at a reasonable price is not enough anymore.
That’s why we have developed 5 beneficial guarantees that will make your experience with our service enjoyable, easy, and safe.

Money-back guarantee

You have to be 100% sure of the quality of your product to give a money-back guarantee. This describes us perfectly. Make sure that this guarantee is totally transparent.

Read more

Zero-plagiarism guarantee

Each paper is composed from scratch, according to your instructions. It is then checked by our plagiarism-detection software. There is no gap where plagiarism could squeeze in.

Read more

Free-revision policy

Thanks to our free revisions, there is no way for you to be unsatisfied. We will work on your paper until you are completely happy with the result.

Read more

Privacy policy

Your email is safe, as we store it according to international data protection rules. Your bank details are secure, as we use only reliable payment systems.

Read more

Fair-cooperation guarantee

By sending us your money, you buy the service we provide. Check out our terms and conditions if you prefer business talks to be laid out in official language.

Read more

Calculate the price of your order

550 words
We'll send you the first draft for approval by September 11, 2018 at 10:52 AM
Total price:
$26
The price is based on these factors:
Academic level
Number of pages
Urgency

Order your essay today and save 10% with the discount code tCPCOVID10