Search icon
Arrow left icon
All Products
Best Sellers
New Releases
Books
Videos
Audiobooks
Learning Hub
Newsletters
Free Learning
Arrow right icon
Data Labeling in Machine Learning with Python

You're reading from  Data Labeling in Machine Learning with Python

Product type Book
Published in Jan 2024
Publisher Packt
ISBN-13 9781804610541
Pages 398 pages
Edition 1st Edition
Languages
Author (1):
Vijaya Kumar Suda Vijaya Kumar Suda
Profile icon Vijaya Kumar Suda

Table of Contents (18) Chapters

Preface Part 1: Labeling Tabular Data
Chapter 1: Exploring Data for Machine Learning Chapter 2: Labeling Data for Classification Chapter 3: Labeling Data for Regression Part 2: Labeling Image Data
Chapter 4: Exploring Image Data Chapter 5: Labeling Image Data Using Rules Chapter 6: Labeling Image Data Using Data Augmentation Part 3: Labeling Text, Audio, and Video Data
Chapter 7: Labeling Text Data Chapter 8: Exploring Video Data Chapter 9: Labeling Video Data Chapter 10: Exploring Audio Data Chapter 11: Labeling Audio Data Chapter 12: Hands-On Exploring Data Labeling Tools Index Other Books You May Enjoy

Using k-means clustering to label regression data

In this section, we are going to use the unsupervised K-means clustering method to label the regression data. We use K-means to cluster data points into groups or clusters based on their similarity.

Once the clustering is done, we can compute the average label value for each cluster by taking the mean of the labeled data points that belong to that cluster. This is because the labeled data points in a cluster are likely to have similar label values since they are similar in terms of their feature values.

Figure 3.6 – Basic k-means clustering with no. of clusters =3

Figure 3.6 – Basic k-means clustering with no. of clusters =3

For example, let’s say we have a dataset of house prices with which we want to predict the price of a house based on features such as size, location, number of rooms, and so on. We have some labeled data points that consist of the features and their corresponding prices, but we also have some unlabeled data points with the same...

lock icon The rest of the chapter is locked
Register for a free Packt account to unlock a world of extra content!
A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.
Unlock this book and the full library FREE for 7 days
Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of
Renews at $15.99/month. Cancel anytime}