← All topics

Learn free · topic 110

Data Labelling

Data Labelling is about tagging something which is not identifiable coming from source. The Labelling can be on known data and on unknown data as well.

A screen shot of several arrows

Description automatically generatedLet’s try to understand the relationship among all these 3 terms i.e., Data Labelling, Data Classification and Data Clustering.

  • Data Clustering group the data based on existing parameter and falls under Unsupervised Learning.
  • Data Labelling is the process of tagging data based on some defined parameter.
  • Data Classification group the data based on manually LABELLED data and falls under Supervised Learning.

Since the era of unstructured dataset, data is coming in many formats. There are specialized tools which can tag or label the unstructured data set so it can become input for Machine Learning (ML) Models. There are also certain libraries used in ML Models to label unstructured data set which can then become input of another ML Model for Supervised Learning.

Finished reading? Test yourself with 10 questions on this topic.

Go to the questions →

From I Am Datapedia! by Mustafa Qizilbash, published here free by the author. Nothing about your reading is stored.