CHF 91.20

Python Data Cleaning Cookbook - Second Edition
Prepare your data for analysis with pandas, NumPy, Matplotlib, scikit-learn, and OpenAI Inglese · Tascabile

Spedizione di solito entro 1 a 2 settimane

Descrizione

Ulteriori informazioni










Learn the intricacies of data description, issue identification, and practical problem-solving, armed with essential techniques and expert tips.Key Features:
- Get to grips with new techniques for data preprocessing and cleaning for machine learning and NLP models
- Use new and updated AI tools and techniques for data cleaning tasks
- Clean, monitor, and validate large data volumes to diagnose problems using cutting-edge methodologies including Machine learning and AIBook Description:
Jumping into data analysis without proper data cleaning will certainly lead to incorrect results. The Python Data Cleaning Cookbook will show you tools and techniques for cleaning and handling data with Python for better outcomes.
Fully updated to the latest version of Python and all relevant tools, this book will teach you how to manipulate and clean data to get it into a useful form. The current edition emphasizes advanced techniques like machine learning and AI-specific approaches and tools to data cleaning along with the conventional ones. The book also delves into tips and techniques to process and clean data for ML, AI and NLP models You will learn how to filter and summarize data to gain insights and better understand what makes sense and what does not, along with discovering how to operate on data to address the issues you've identified. Next, you'll cover recipes for using supervised learning and Naive Bayes analysis to identify unexpected values and classification errors and generate visualizations for exploratory data analysis (EDA) to identify unexpected values. Finally, you'll build functions and classes that you can reuse without modification when you have new data.
By the end of this Data Cleaning book, you'll know how to clean data and diagnose problems within it.What You Will Learn:
- Using OpenAI tools for various data cleaning tasks
- Produce summaries of the attributes of datasets, columns, and rows
- Anticipating Data Cleaning Issues when Importing Tabular Data into Pandas
- Apply validation techniques for imported tabular data
- Improve your productivity in Python pandas by using method chaining
- Recognize and resolve common issues like dates and IDs
- Set up indexes to streamline data issue identification
- Use data cleaning to prepare your data for ML and AI modelsWho this book is for:
This book is for anyone looking for ways to handle messy, duplicate, and poor data using different Python tools and techniques. The book takes a recipe-based approach to help you to learn how to clean and manage data with practical examples.
Working knowledge of Python programming is all you need to get the most out of the book.


Info autore










Michael Walker has worked as a data analyst for over 30 years at a variety of educational institutions. He has also taught data science, research methods, statistics, and computer programming to undergraduates since 2006. He is currently the Chief Information Officer at College Unbound in Providence, Rhode Island.


Dettagli sul prodotto

Autori Michael Walker
Editore Packt Publishing
 
Contenuto Libro
Forma del prodotto Tascabile
Data pubblicazione 31.05.2024
Categoria Guide e manuali
Scienze naturali, medicina, informatica, tecnica > Informatica, EDP > Informatica
 
EAN 9781803239873
ISBN 978-1-80323-987-3
Numero di pagine 486
Dimensioni (della confezione) 19.1 x 23.5 x 2.6 cm
Peso (della confezione) 899 g
 
Categorie Pandas
numpy
Data Processing
 

Recensioni dei clienti

Per questo articolo non c'è ancora nessuna recensione. Scrivi la prima recensione e aiuta gli altri utenti a scegliere.

Scrivi una recensione

Top o flop? Scrivi la tua recensione.

Per i messaggi a CeDe.ch si prega di utilizzare il modulo di contatto.

I campi contrassegnati da * sono obbligatori.

Inviando questo modulo si accetta la nostra dichiarazione protezione dati.