
Python Data Cleaning Cookbook - Second Edition: Prepare your data for analysis with pandas, NumPy, Matplotlib, scikit-le, (Paperback)
Key item features
- Python Data Cleaning Cookbook - Second Edition: Prepare your data for analysis with pandas, NumPy, Matplotlib, scikit-le, (Paperback)
- Author: Packt Publishing
- ISBN: 9781803239873
- Format: Paperback
- Publication Date: 2024-05-31
- Page Count: 486
Specs
- Book formatPaperback
- Fiction/nonfictionNon-Fiction
- GenreComputing & Internet
- Pages486
- SubgenreData Science
- Series titleNo Series
- Free shipping
Free 90-day returns
How do you want your item?
Get free delivery, shipping and more*
About this item
Product details
Learn the intricacies of data description, issue identification, and practical problem-solving, armed with essential techniques and expert tips.
Key Features:- Get to grips with new techniques for data preprocessing and cleaning for machine learning and NLP models
- Use new and updated AI tools and techniques for data cleaning tasks
- Clean, monitor, and validate large data volumes to diagnose problems using cutting-edge methodologies including Machine learning and AI
Book Description:Jumping into data analysis without proper data cleaning will certainly lead to incorrect results. The Python Data Cleaning Cookbook will show you tools and techniques for cleaning and handling data with Python for better outcomes.
Fully updated to the latest version of Python and all relevant tools, this book will teach you how to manipulate and clean data to get it into a useful form. The current edition emphasizes advanced techniques like machine learning and AI-specific approaches and tools to data cleaning along with the conventional ones. The book also delves into tips and techniques to process and clean data for ML, AI and NLP models You will learn how to filter and summarize data to gain insights and better understand what makes sense and what does not, along with discovering how to operate on data to address the issues you've identified. Next, you'll cover recipes for using supervised learning and Naive Bayes analysis to identify unexpected values and classification errors and generate visualizations for exploratory data analysis (EDA) to identify unexpected values. Finally, you'll build functions and classes that you can reuse without modification when you have new data.
By the end of this Data Cleaning book, you'll know how to clean data and diagnose problems within it.
What You Will Learn:- Using OpenAI tools for various data cleaning tasks
- Produce summaries of the attributes of datasets, columns, and rows
- Anticipating Data Cleaning Issues when Importing Tabular Data into Pandas
- Apply validation techniques for imported tabular data
- Improve your productivity in Python pandas by using method chaining
- Recognize and resolve common issues like dates and IDs
- Set up indexes to streamline data issue identification
- Use data cleaning to prepare your data for ML and AI models
Who this book is for:This book is for anyone looking for ways to handle messy, duplicate, and poor data using different Python tools and techniques. The book takes a recipe-based approach to help you to learn how to clean and manage data with practical examples.
Working knowledge of Python programming is all you need to get the most out of the book.
- Python Data Cleaning Cookbook - Second Edition: Prepare your data for analysis with pandas, NumPy, Matplotlib, scikit-le, (Paperback)
- Author: Packt Publishing
- ISBN: 9781803239873
- Format: Paperback
- Publication Date: 2024-05-31
- Page Count: 486
Specifications
Book format
Fiction/nonfiction
Genre
Pages
Warranty
Warranty information
Similar items you might like
Based on what customers bought
Scientific Computing with Python: Mastering Numpy and Scipy, (Paperback) $19.95
$1995current price $19.95Scientific Computing with Python: Mastering Numpy and Scipy, (Paperback)
Bioinformatics with Python Cookbook - Fourth Edition: Solve advanced computational biology problems and build production, (Paperback) $44.99
$4499current price $44.99Bioinformatics with Python Cookbook - Fourth Edition: Solve advanced computational biology problems and build production, (Paperback)
Cognitive Technologies Python for Natural Language Processing: Programming with Numpy, Scikit-Learn, Keras, and Pytorch, (Hardcover) $46.20
$4620current price $46.20Cognitive Technologies Python for Natural Language Processing: Programming with Numpy, Scikit-Learn, Keras, and Pytorch, (Hardcover)
Pandas Cookbook - Third Edition: Practical recipes for scientific computing, time series, and exploratory data analysis , (Paperback) $39.99
$3999current price $39.99Pandas Cookbook - Third Edition: Practical recipes for scientific computing, time series, and exploratory data analysis , (Paperback)
Self-Learning Management Python Essentials You Always Wanted to Know: Beginner's Guide to Python Programming, Data Structures, Data Analytic, (Paperback) $31.99
$3199current price $31.99Self-Learning Management Python Essentials You Always Wanted to Know: Beginner's Guide to Python Programming, Data Structures, Data Analytic, (Paperback)
Python Data Cleaning Cookbook: Modern techniques and Python tools to detect and remove dirty data and extract key insights (Paperback) $40.10
$4010current price $40.10Python Data Cleaning Cookbook: Modern techniques and Python tools to detect and remove dirty data and extract key insights (Paperback)
Time Series Indexing: Implement iSAX in Python to index time series with confidence (Paperback) $49.99
$4999current price $49.99Time Series Indexing: Implement iSAX in Python to index time series with confidence (Paperback)
Polars Cookbook: Over 60 practical recipes to transform, manipulate, and analyze your data using Python Polars 1.x, (Paperback) $47.42
$4742current price $47.42Polars Cookbook: Over 60 practical recipes to transform, manipulate, and analyze your data using Python Polars 1.x, (Paperback)
Hands-On Data Preprocessing in Python: Learn how to effectively prepare data for successful data analytics (Paperback) $51.72
$5172current price $51.72Hands-On Data Preprocessing in Python: Learn how to effectively prepare data for successful data analytics (Paperback)
Numerical Python: Scientific Computing and Data Science Applications with Numpy, Scipy and Matplotlib, (Paperback) $29.24
$2924current price $29.24Numerical Python: Scientific Computing and Data Science Applications with Numpy, Scipy and Matplotlib, (Paperback)
Python for Data Analysis: A Step-By-Step Guide to Master the Basics of Data Science and Analysis in Python Using Pandas, $37.13 Was $41.54
$3713current price $37.13, Was $41.54$41.54Python for Data Analysis: A Step-By-Step Guide to Master the Basics of Data Science and Analysis in Python Using Pandas,
Computational Framework for the Finite Element Method in MATLAB(R) and Python, (Paperback) $46.49
$4649current price $46.49Computational Framework for the Finite Element Method in MATLAB(R) and Python, (Paperback)
Practical Python Data Visualization: A Fast Track Approach to Learning Data Visualization with Python, (Paperback) $47.16
$4716current price $47.16Practical Python Data Visualization: A Fast Track Approach to Learning Data Visualization with Python, (Paperback)
Learn Model Context Protocol with Python: Build agentic systems in Python with the new standard for AI capabilities, (Paperback) $36.99
$3699current price $36.99Learn Model Context Protocol with Python: Build agentic systems in Python with the new standard for AI capabilities, (Paperback)
Handbook of Regression Modeling in People Analytics: With Examples in R and Python, (Paperback) $65.99
$6599current price $65.99Handbook of Regression Modeling in People Analytics: With Examples in R and Python, (Paperback)
Deep Learning for Time Series Cookbook: Use PyTorch and Python recipes for forecasting, classification, and anomaly dete, (Paperback) $39.99
$3999current price $39.99Deep Learning for Time Series Cookbook: Use PyTorch and Python recipes for forecasting, classification, and anomaly dete, (Paperback)
Data Analysis from Scratch with Python Bundle: Basic Data Analysis and Time Series Analysis in Finance using Python, (Hardcover) $49.99
$4999current price $49.99Data Analysis from Scratch with Python Bundle: Basic Data Analysis and Time Series Analysis in Finance using Python, (Hardcover)
Python for Data Science: The Ultimate Step-by-Step Guide to Python Programming. Discover How to Master Big Data Analysis, (Paperback) $19.75
$1975current price $19.75Python for Data Science: The Ultimate Step-by-Step Guide to Python Programming. Discover How to Master Big Data Analysis, (Paperback)
Python Simplified with Generative AI: Hands-on Python development with GenAI tools integrating data science and web inte, (Paperback) $46.65
$4665current price $46.65Python Simplified with Generative AI: Hands-on Python development with GenAI tools integrating data science and web inte, (Paperback)
Pandas Cookbook: Recipes for Scientific Computing, Time Series Analysis and Data Visualization using Python (Paperback) $40.10 Was $49.99
$4010current price $40.10, Was $49.99$49.99Pandas Cookbook: Recipes for Scientific Computing, Time Series Analysis and Data Visualization using Python (Paperback)
