Inicio > > Bases de datos > Building ETL Pipelines with Python
Building ETL Pipelines with Python

Building ETL Pipelines with Python

Brij Kishore Pandey / Emily Ro Schoof

55,04 €
IVA incluido
Disponible
Editorial:
Packt Publishing
Año de edición:
2023
Materia
Bases de datos
ISBN:
9781804615256
55,04 €
IVA incluido
Disponible

Selecciona una librería:

  • Librería 7artes
  • Donde los libros
  • Librería Elías (Asturias)
  • Librería Kolima (Madrid)
  • Librería Proteo (Málaga)

Develop production-ready ETL pipelines by leveraging Python libraries and deploying them for suitable use casesKey Features:Understand how to set up a Python virtual environment with PyCharmLearn functional and object-oriented approaches to create ETL pipelinesCreate robust CI/CD processes for ETL pipelinesPurchase of the print or Kindle book includes a free PDF eBookBook Description:Modern extract, transform, and load (ETL) pipelines for data engineering have favored the Python language for its broad range of uses and a large assortment of tools, applications, and open source components. With its simplicity and extensive library support, Python has emerged as the undisputed choice for data processing.In this book, you’ll walk through the end-to-end process of ETL data pipeline development, starting with an introduction to the fundamentals of data pipelines and establishing a Python development environment to create pipelines. Once you’ve explored the ETL pipeline design principles and ET development process, you’ll be equipped to design custom ETL pipelines. Next, you’ll get to grips with the steps in the ETL process, which involves extracting valuable data; performing transformations, through cleaning, manipulation, and ensuring data integrity; and ultimately loading the processed data into storage systems. You’ll also review several ETL modules in Python, comparing their pros and cons when building data pipelines and leveraging cloud tools, such as AWS, to create scalable data pipelines. Lastly, you’ll learn about the concept of test-driven development for ETL pipelines to ensure safe deployments.By the end of this book, you’ll have worked on several hands-on examples to create high-performance ETL pipelines to develop robust, scalable, and resilient environments using Python.What You Will Learn:Explore the available libraries and tools to create ETL pipelines using PythonWrite clean and resilient ETL code in Python that can be extended and easily scaledUnderstand the best practices and design principles for creating ETL pipelinesOrchestrate the ETL process and scale the ETL pipeline effectivelyDiscover tools and services available in AWS for ETL pipelinesUnderstand different testing strategies and implement them with the ETL processWho this book is for:If you are a data engineer or software professional looking to create enterprise-level ETL pipelines using Python, this book is for you. Fundamental knowledge of Python is a prerequisite.

Artículos relacionados

  • Strategic Advancements in Utilizing Data Mining and Warehousing Technologies
    Organizations rely on data mining and warehousing technologies to store, integrate, query, and analyze essential data. Strategic Advancements in Utilizing Data Mining and Warehousing Technologies: New Concepts and Developments discusses developments in data mining and warehousing as well as techniques for successful implementation. Contributions investigate theoretical queries ...
    Disponible

    236,66 €

  • Online Instruments, Data Collection, and Electronic Measurements
    One of the infinite rewards to continuously advancing technology is an increased ease and precision in organizational techniques. Online data collection and online instruments are vital ways to electronically measure and assess organizational areas relevant to management, leadership, and human research development. Online Instruments, Data Collection, and Electronic Measurement...
    Disponible

    229,95 €

  • Exploring Advances in Interdisciplinary Data Mining and Analytics
    Data mining is still a relatively young field, expanding at the rate of technology while advancing tools and techniques for gaining knowledge, finding patterns, and managing databases. Exploring Advances in Interdisciplinary Data Mining and Analytics: New Trends is an updated look at the state of technology in the field of data mining and analytics. As processor speeds, databas...
    Disponible

    255,93 €

  • Research and Trends in Data Mining Technologies and Applications
    David Taniar
    ...
    Disponible

    88,74 €

  • Knowledge Discovery Practices and Emerging Applications of Data Mining
    Recent developments have drastically increased the volume and complexity of data available to be mined, leading researchers to explore new ways to glean non-trivial data automatically. Knowledge Discovery Practices and Emerging Applications of Data Mining: Trends and New Domains introduces the reader to recent research activities in the field of data mining. This book covers as...
    Disponible

    236,54 €

  • Developing Metadata Application Profiles
    The prevalence of data science has grown exponentially in recent years. Increases in data exchange have created the need for standards and formats on handling data from different sources. Developing Metadata Application Profiles is an innovative reference source that discusses the latest trends and techniques for effectively managing and exchanging metadata. Including a range o...
    Disponible

    223,01 €