Hands-On Data Preprocessing in Python

By Roy Jafari
    Advance your knowledge in tech with a Packt subscription

  • Instant online access to over 7,500+ books and videos
  • Constantly updated with 100+ new titles each month
  • Breadth and depth in over 1,000+ technologies

About this book

Data preprocessing is the first step in data visualization, data analytics, and machine learning, where data is prepared for analytics functions to get the best possible insights. Around 90% of the time spent on data analytics, data visualization, and machine learning projects is dedicated to performing data preprocessing.

This book will equip you with optimum data preprocessing techniques from multiple perspectives. You'll learn different technical and analytical aspects of data preprocessing - data collection, data cleaning, data integration, data reduction, and data transformation - and get to grips with implementing them using the open-source Python programming environment. The book will provide a comprehensive articulation of data preprocessing, its whys and hows, and help you identify analytics opportunities where data analytics could lead to more effective decision making. It also demonstrates the role of data management systems and technologies for effective analytics and how to create queries to pull data from relational databases.

By the end of this Python data preprocessing book, you'll be able to use Python to read, manipulate, and analyze data, perform data cleaning, integration, reduction techniques, and handle outliers or missing values to implement the appropriate data transformation method.

Publication date:
December 2021
Publisher
Packt
Pages
554
ISBN
9781801072137

About the Author

  • Roy Jafari

    Roy Jafari, Ph.D. is an Assistant Professor of Business Analytics at the University of Redlands.

    Roy has taught and developed advanced college-level courses that cover data cleaning, decision making, data science, data mining, machine learning, and optimization.

    Roy’s style of teaching is hands-on and he believes the best way to learn is to learn by doing. He likes to use the active learning teaching philosophy and Readers will get to experience active learning in this book.

    Roy, who is known as Ruholla Jafari Marandi in the research community, has authored scientific research publications. In his research, Roy has contributed new data mining and decision-making algorithms. He has also shown various applications of data mining and machine learning and provided steps for more fair and equitable applications of machine learning.

    Roy believes successful data preprocessing happens only when one is equipped with the most efficient tools, has the appropriate understanding of data analytic goals, and knows the steps of data preprocessing, and can compare a variety of methods. This belief has shaped the structure his upcoming book Hands-on Data Preprocessing in Python.

    Browse publications by this author
Hands-On Data Preprocessing in Python
Unlock this book and the full library for FREE
Start free trial