The Data Lakehouse

The Data Lakehouse

Bill Inmon / Dave Rapien / Valerie Bartelt

42,96 €
IVA incluido
Disponible
Editorial:
Technics Publications
Año de edición:
2023
ISBN:
9781634621571
42,96 €
IVA incluido
Disponible

Selecciona una librería:

  • Librería Samer Atenea
  • Librería Aciertas (Toledo)
  • Kálamo Books
  • Librería Perelló (Valencia)
  • Librería Elías (Asturias)
  • Donde los libros
  • Librería Kolima (Madrid)
  • Librería Proteo (Málaga)

The data lakehouse is the next generation of the data warehouse and data lake, designed to meet today’s complex and ever-changing modern information systems. This book shows you how to construct your data lakehouse as the foundation for your artificial intelligence (AI), machine learning (ML), and data mesh initiatives. Know the pitfalls and techniques for maximizing business value of your data lakehouse.In addition, be able to explain the core characteristics and critical success factors of a data lakehouse. By reviewing entry errors, key incompatibility, and ensuring good documentation, we can improve the data quality and believability of your lakehouse. Evaluate criteria for data quality, including accuracy, completeness, reliability, relevance, and timeliness. Understand the different types of storage for the lakehouse, including the under-utilized yet extremely valuable bulk storage. There are three data types in the data lakehouse (structured, textual, and analog/ IoT), and for each, learn how to build a robust foundation for artificial intelligence (AI), machine learning (ML), and data mesh. Leverage data models for structured data, ontologies and taxonomies for textual data, and distillation algorithms for analog/IoT data. Learn how to abstract these data types to accommodate future requirements and simplify data lineage. Apply Extract, Transform, and Load (ETL) to create a structure that returns the answers to business problems. The end result is a data lakehouse that meets our needs. Speaking of human needs, learn Maslow’s Hierarchy of Data Lakehouse Needs. Next explore data integration geared for Al, ML, and data mesh. Then deep dive with us into all of the varieties of analytics within the lakehouse, including structured, textual, and analog analytics. Witness how descriptive data, data catalog, and metadata can increase the value of the lakehouse. We conclude with a detailed evolution of data architecture, from magnetic tape to the data lakehouse as a bedrock foundation for AI, ML, and data mesh.

Artículos relacionados

  • Exploring Advances in Interdisciplinary Data Mining and Analytics
    Data mining is still a relatively young field, expanding at the rate of technology while advancing tools and techniques for gaining knowledge, finding patterns, and managing databases. Exploring Advances in Interdisciplinary Data Mining and Analytics: New Trends is an updated look at the state of technology in the field of data mining and analytics. As processor speeds, databas...
  • Knowledge Discovery Practices and Emerging Applications of Data Mining
    Recent developments have drastically increased the volume and complexity of data available to be mined, leading researchers to explore new ways to glean non-trivial data automatically. Knowledge Discovery Practices and Emerging Applications of Data Mining: Trends and New Domains introduces the reader to recent research activities in the field of data mining. This book covers as...
  • Research and Trends in Data Mining Technologies and Applications
    David Taniar
    ...
  • Developing Metadata Application Profiles
    The prevalence of data science has grown exponentially in recent years. Increases in data exchange have created the need for standards and formats on handling data from different sources. Developing Metadata Application Profiles is an innovative reference source that discusses the latest trends and techniques for effectively managing and exchanging metadata. Including a range o...
  • Modern Technologies for Big Data Classification and Clustering
    Data has increased due to the growing use of web applications and communication devices. It is necessary to develop new techniques of managing data in order to ensure adequate usage. Modern Technologies for Big Data Classification and Clustering is an essential reference source for the latest scholarly research on handling large data sets with conventional data mining and provi...
  • 90 Gelöste Fälle zu Zeitintelligenz in der DAX-Sprache
    Ramón Javier Castro Amador
    Dieser Ratgeber ist rein praktisch ausgerichtet, so dass Sie den gesamten DAX-Code in dieser Publikation anhand einer zum Download verfügbaren .pbix-Datei testen können.'90 gelöste Fälle zu Zeitintelligenz in DAX' ist ein Ratgeber für Benutzer von Microsoft Power BI, der Lösungen für sehr häufige praktische Fälle in Zeitintelligenzmodellen in der Sprache DAX bietet.Um das Verst...
    Disponible

    16,15 €