Data Quality - Data (cleaning|scrubbing|wrangling)

1 - About

Data cleansing or data scrubbing is the act of:

corrupt or inaccurate records from a record set, table, or database.

Used mainly in databases, the term refers to identifying incomplete, incorrect, inaccurate, irrelevant etc. parts of the data and then replacing, modifying or deleting this dirty data.

After cleansing, a data set will be consistent with other similar data sets in the system. The inconsistencies detected or removed may have been originally caused by different data dictionary definitions of similar entities in different stores, may have been caused by user entry errors, or may have been corrupted in transmission or storage.

3 - Tool

4 - Documentation / Reference

  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at del.icio.us
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at Digg
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at Ask
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at Google
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at StumbleUpon
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at Technorati
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at Live Bookmarks
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at Yahoo! Myweb
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at Facebook
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at Yahoo! Bookmarks
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at Twitter
  • Bookmark "Data Quality - Data (cleaning|scrubbing|wrangling)" at myAOL
data_quality/data_cleansing.txt ยท Last modified: 2016/11/16 10:50 by gerardnico