Data Preprocessing/Exercise Sheet 2

May 3, 2020 Leave a Comment Data Mining, Data Mining

Theory:
Data Preprocessing in the Data Mining Process:

The data mining/KDD process
Why data preprocessing?

Issues in Data Preprocessing:

Data Cleaning
Data Transformation
Variable Construction
Data Reduction and Discretization
Data Integration

The data mining/KDD Process:
Understanding customer: 10%-20%
Understanding data:20-30
Prepare data: 40-70%
Build Models: 10-20%
Evaluate models: 10%-20%
Take action:10%20%

Why data mining?

Real – world data is dirty
Low data quality anyway a huge problem in data mining
Garbage in,garbage out
Different methods, different requirements

R Working Codes for data mining:

R code is case sensitive:
I am doing it from professors sheet.

dim means dimension

This line i could not make work:

hist(Ozone,breaks=25,ylim=(c(0,45)),main=”Original data”)

And another question how the imputation works

Exercise 2 (K)= I have to find the answers

Exercise 3: Answer:

clothing=read.csv(file="F:/desktop and documents/Desktop/dataminingdata/clothing_store.txt")

It would be a great help, if you support by sharing :)

Author: zakilive

Leave a Reply Cancel reply

Design by ThemesDNA.com

Privacy Overview

This website uses cookies to improve your experience while you navigate through the website. Out of these, the cookies that are categorized as necessary are stored on your browser as they are essential for the working of basic functionalities of the website. We also use third-party cookies that help us analyze and understand how you use this website. These cookies will be stored in your browser only with your consent. You also have the option to opt-out of these cookies. But opting out of some of these cookies may affect your browsing experience.

Necessary

Always Enabled

Necessary cookies are absolutely essential for the website to function properly. These cookies ensure basic functionalities and security features of the website, anonymously.

Cookie	Duration	Description
cookielawinfo-checkbox-analytics	11 months	This cookie is set by GDPR Cookie Consent plugin. The cookie is used to store the user consent for the cookies in the category "Analytics".
cookielawinfo-checkbox-functional	11 months	The cookie is set by GDPR cookie consent to record the user consent for the cookies in the category "Functional".
cookielawinfo-checkbox-necessary	11 months	This cookie is set by GDPR Cookie Consent plugin. The cookies is used to store the user consent for the cookies in the category "Necessary".
cookielawinfo-checkbox-others	11 months	This cookie is set by GDPR Cookie Consent plugin. The cookie is used to store the user consent for the cookies in the category "Other.
cookielawinfo-checkbox-performance	11 months	This cookie is set by GDPR Cookie Consent plugin. The cookie is used to store the user consent for the cookies in the category "Performance".
viewed_cookie_policy	11 months	The cookie is set by the GDPR Cookie Consent plugin and is used to store whether or not user has consented to the use of cookies. It does not store any personal data.

Functional

Performance

Analytics

Others

Data Preprocessing/Exercise Sheet 2

Leave a Reply Cancel reply

Recent Posts

Archives

Subscribe to my YouTube channel !

Categories

Recent Comments

Leave a Reply Cancel reply

Recent Posts

Tags

Archives

Subscribe to my YouTube channel !

Categories

Recent Comments