Python data cleaning cookbook: prepare your data for analysis with pandas, NumPy, Matplotlib, scikit-learn and OpenAI

Jumping into data analysis without proper data cleaning will certainly lead to incorrect results. The Python Data Cleaning Cookbook - Second Edition will show you tools and techniques for cleaning and handling data with Python for better outcomes. Fully updated to the latest version of Python and al...

Ausführliche Beschreibung

Gespeichert in:
Bibliographische Detailangaben
Beteilige Person: Walker, Michael (VerfasserIn)
Format: Elektronisch E-Book
Sprache:Englisch
Veröffentlicht: Birmingham, UK Packt Publishing Ltd. 2024
Ausgabe:Second edition.
Schriftenreihe:Expert insight
Schlagwörter:
Links:https://learning.oreilly.com/library/view/-/9781803239873/?ar
Zusammenfassung:Jumping into data analysis without proper data cleaning will certainly lead to incorrect results. The Python Data Cleaning Cookbook - Second Edition will show you tools and techniques for cleaning and handling data with Python for better outcomes. Fully updated to the latest version of Python and all relevant tools, this book will teach you how to manipulate and clean data to get it into a useful form. he current edition focuses on advanced techniques like machine learning and AI-specific approaches and tools for data cleaning along with the conventional ones. The book also delves into tips and techniques to process and clean data for ML, AI, and NLP models. You will learn how to filter and summarize data to gain insights and better understand what makes sense and what does not, along with discovering how to operate on data to address the issues you've identified. Next, you'll cover recipes for using supervised learning and Naive Bayes analysis to identify unexpected values and classification errors and generate visualizations for exploratory data analysis (EDA) to identify unexpected values. Finally, you'll build functions and classes that you can reuse without modification when you have new data. By the end of this Data Cleaning book, you'll know how to clean data and diagnose problems within it.
Beschreibung:Includes index
Umfang:1 Online-Ressource (486 Seiten) illustrations
ISBN:9781803239873