Learning Pandas 2, Second Edition

Learning Pandas 2, Second Edition

Matthew Rosch

61,65 €
IVA incluido
Disponible
Editorial:
GitforGits
Año de edición:
2026
ISBN:
9789349174665
61,65 €
IVA incluido
Disponible

Selecciona una librería:

  • Librería Samer Atenea
  • Kálamo Books
  • Librería Elías (Asturias)
  • Librería Kolima (Madrid)
  • Librería Proteo (Málaga)

This book has been updated with Pandas 2.3, and it’s exactly what ML engineers, data scientists and data engineers have been waiting for. It’s a hands-on desk guide that’s full of solutions, and it’s the most up-to-date, production-ready book to the most widely used data manipulation library in the Python ecosystem.This book covers all the big changes in Pandas 2.3, like Copy-on-Write semantics, PyArrow-backed types that save over 50% memory, the new default StringDtype, and the deprecated frequency aliases that are messing up time series pipelines everywhere. All the chapters are based on one growing application using a real Customer Churn dataset, so every technique is put into a context where you can trace it and use it in production.Once you’ve got the hang of pandas, you will be exploring deep into feature engineering with feature_engine and scikit-learn’s set_output API, dealing with class imbalance with SMOTE and ADASYN, and doing distributed computing with Dask, as well as JIT-compiled custom functions with Numba and JAX. On top of that, you’ll be able to handle full NLP pipelines from TF-IDF to LDA topic modelling, and geospatial analysis with GeoPandas.It doesn’t matter if you’re building ML pipelines, scaling data infrastructure, or connecting pandas to TensorFlow, PyTorch, or JAX, this book will give you the practical depth and modern patterns to do it correctly on pandas 2.3 today, and stay forward-compatible with pandas 3.0 tomorrow.Key FeaturesBuild memory-efficient pipelines using PyArrow backends and targeted dtype choices.Write Copy-on-Write-safe assignment patterns that work on pandas 2.3 and 3.0.Engineer rich ML features using ratios, bins, group statistics, and interaction terms.Handle class imbalance with SMOTE, ADASYN, and quantified pandas-based profiling.Scale datasets beyond RAM using Dask lazy evaluation and distributed cluster computing.Accelerate custom scoring functions with Numba JIT and JAX-compiled batch operations.Extract sentiment, topics, and clusters from raw text using TF-IDF and LDA pipelines.Perform spatial joins, buffer analysis, and geocoding with GeoPandas and geopy.Preserve named DataFrames throughout sklearn Pipelines using the set_output API.Migrate confidently from legacy pandas patterns to pandas 2.3 production standards.Table of ContentGetting Started with Pandas 2.3Data Read, Storage, and File FormatsIndexing and Selecting DataData Manipulation and TransformationTime Series and DateTime OperationsPerformance Optimization and ScalingMachine Learning with Pandas 2.3Text Mining and NLPGeospatial Data Analysis

Artículos relacionados

  • Exploring Advances in Interdisciplinary Data Mining and Analytics
    Data mining is still a relatively young field, expanding at the rate of technology while advancing tools and techniques for gaining knowledge, finding patterns, and managing databases. Exploring Advances in Interdisciplinary Data Mining and Analytics: New Trends is an updated look at the state of technology in the field of data mining and analytics. As processor speeds, databas...
  • Knowledge Discovery Practices and Emerging Applications of Data Mining
    Recent developments have drastically increased the volume and complexity of data available to be mined, leading researchers to explore new ways to glean non-trivial data automatically. Knowledge Discovery Practices and Emerging Applications of Data Mining: Trends and New Domains introduces the reader to recent research activities in the field of data mining. This book covers as...
  • Research and Trends in Data Mining Technologies and Applications
    David Taniar
    ...
  • Developing Metadata Application Profiles
    The prevalence of data science has grown exponentially in recent years. Increases in data exchange have created the need for standards and formats on handling data from different sources. Developing Metadata Application Profiles is an innovative reference source that discusses the latest trends and techniques for effectively managing and exchanging metadata. Including a range o...
  • Modern Technologies for Big Data Classification and Clustering
    Data has increased due to the growing use of web applications and communication devices. It is necessary to develop new techniques of managing data in order to ensure adequate usage. Modern Technologies for Big Data Classification and Clustering is an essential reference source for the latest scholarly research on handling large data sets with conventional data mining and provi...
  • 90 Gelöste Fälle zu Zeitintelligenz in der DAX-Sprache
    Ramón Javier Castro Amador
    Dieser Ratgeber ist rein praktisch ausgerichtet, so dass Sie den gesamten DAX-Code in dieser Publikation anhand einer zum Download verfügbaren .pbix-Datei testen können.'90 gelöste Fälle zu Zeitintelligenz in DAX' ist ein Ratgeber für Benutzer von Microsoft Power BI, der Lösungen für sehr häufige praktische Fälle in Zeitintelligenzmodellen in der Sprache DAX bietet.Um das Verst...
    Disponible

    16,15 €

Otros libros del autor

  • Learning PyTorch 2.0, Second Edition
    Matthew Rosch
    'Learning PyTorch 2.0, Second Edition' is a fast-learning, hands-on book that emphasizes practical PyTorch scripting and efficient model development using PyTorch 2.3 and CUDA 12. This edition is centered on practical applications and presents a concise methodology for attaining proficiency in the most recent features of PyTorch. The book presents a practical program based on t...
    Disponible

    65,56 €

  • PyTorch Cookbook
    Matthew Rosch
    Starting a PyTorch Developer and Deep Learning Engineer career? Check out this ’PyTorch Cookbook,’ a comprehensive guide with essential recipes and solutions for PyTorch and the ecosystem. The book covers PyTorch deep learning development from beginner to expert in well-written chapters.The book simplifies neural networks, training, optimization, and deployment strategies chapt...
    Disponible

    62,95 €

  • Learning PyTorch 2.0
    Matthew Rosch
    This book is a comprehensive guide to understanding and utilizing PyTorch 2.0 for deep learning applications. It starts with an introduction to PyTorch, its various advantages over other deep learning frameworks, and its blend with CUDA for GPU acceleration. We delve into the heart of PyTorch - tensors, learning their different types, properties, and operations. Through step-by...
    Disponible

    52,42 €

  • Learning Pandas 2.0
    Matthew Rosch
    Mastering Data Wrangling and Analysis for Modern Data Science'Learning Pandas 2.0' is an essential guide for anyone looking to harness the power of Python’s premier data manipulation library. With this comprehensive resource, you will not only master core Pandas 2.0 concepts but also learn how to employ its advanced features to perform efficient data manipulation and analysis.T...
    Disponible

    52,72 €