← All topics

Learn free · topic 116

Data Identification

Data Identification is the pre-requisite for Data Anonymization. How?

Diagram

Description automatically generatedData Identification is opposite to Data Anonymization, in other words, first one must identify data before conducting Data Anonymization.

Data Identification is about those set of details which assist in identifying an entity or individual. These details can be direct details like name, contact number, address, credit card number etc. or can be indirect details like unique identification number which assist to connect with other datasets to identify an entity or individual.

Data Identification is a very crucial exercise in data migration project where the process is to first conduct data discovery then data identification and then data migration can take place.

Data Identification can be done in all kinds of datasets i.e., structure, semi or unstructured but recommended practice is to do it in structured datasets.

Other Benefits

  • Data dictionaries and data glossaries are also a good source for data identification.
  • Data Anonymization becomes easy when Metadata Management is in place.
  • Data Lineage also make life easy, to track where identified element has come from or where it will flow to.
  • Data Identification also assists in conducting better Data Cleansing, Data Wrangling etc.
  • With Data Identification, Data Quality implementation becomes highly achievable.
  • While defining Golden Records during Master Data Management or defining Reference data, Data Identification also plays a vital role.

Finished reading? Test yourself with 10 questions on this topic.

Go to the questions →

From I Am Datapedia! by Mustafa Qizilbash, published here free by the author. Nothing about your reading is stored.