Inicio > > Bases de datos > Diseño y teoría de bases de datos > Building Trusted Data Platforms with Azure Databricks and GenAI - Second Edition
Building Trusted Data Platforms with Azure Databricks and GenAI - Second Edition

Building Trusted Data Platforms with Azure Databricks and GenAI - Second Edition

Manoj Kukreja

87,01 €
IVA incluido
Disponible
Editorial:
Packt Publishing
Año de edición:
2026
Materia
Diseño y teoría de bases de datos
ISBN:
9781806679775
87,01 €
IVA incluido
Disponible
Añadir a favoritos

A practical guide to building a modern, GenAI-powered data platform with a Lakehouse foundation, covering MDM, data mesh, AI enablement, streaming pipelines, observability, and cloud-driven architectures for trusted analytics.Key Features:- Discover characteristics of future-ready platforms - data mesh, automation, & observability- Design trustworthy data products with contracts, federated governance, and decentralized ownership- Understand how GenAI accelerates Lakehouse development and enables self-service analyticsBook Description:Discover the defining hallmarks of future-ready data platforms, including data mesh architectures, intelligent automation, and end-to-end data observability. Learn how to design and deliver trusted data products through data contracts, federated governance, decentralized domain ownership, and endorsed datasets. The book explores modern Lakehouse patterns with a strong focus on the medallion architecture, explaining how bronze, silver, and gold layers transform raw data into analytics-ready assets governed through Unity Catalog. You’ll gain practical guidance on MDM linkages, survivorship rules, and entity resolution to ensure consistent master data across domains. It also covers real-time and streaming pipelines that integrate seamlessly with the Lakehouse. We focus on self-service analytics, showing how governed data products let business users explore, analyze, and derive insights independently with confidence. Finally, understand how GenAI accelerates platform development through automated code generation using tools like Claude Code and Databricks Genie Code, enabling faster pipeline creation, governance, and analytics delivery.What You Will Learn:- Future-ready platforms: data mesh, automation, observability- Design trusted data products with contracts and governance- Build Lakehouses with medallion architecture: bronze, silver, gold- Apply Unity Catalog for governance and endorsed datasets- Implement MDM using linkages, survivorship, and entity resolution- Develop real-time and streaming pipelines at scale- Enable governed self-service analytics for business users- Use GenAI to generate code with Claude and Databricks Genie Who this book is for:This book is crafted for aspiring data and AI/ML architects, engineers and analysts starting their data engineering journey and seeking a practical, hands-on guide to building scalable, cloud-driven data platforms. It’s ideal for professionals familiar with PySpark who want to design modern Lakehouse architectures using Delta Lake, while learning MDM, data mesh, AI enablement, streaming pipelines, automation, and data observability. A working knowledge of Python, Spark, and SQL is expected.Table of Contents- The Story of Data Engineering and Analytics- Discovering Storage and Compute in Lakehouses- Data Engineering on Microsoft Azure- Designing Future Data Platforms- Databricks, Medallion Architecture & Delta Lake- Understanding Modern Data Pipelines- Data Collection Stage - The Bronze Layer- Data Curation Stage - The Silver Layer- Data Aggregation Stage - The Gold Layer- Next-Gen Data Analytics with Generative AI- Data Observability- Data Governance

Artículos relacionados

  • Exploring Advances in Interdisciplinary Data Mining and Analytics
    Data mining is still a relatively young field, expanding at the rate of technology while advancing tools and techniques for gaining knowledge, finding patterns, and managing databases. Exploring Advances in Interdisciplinary Data Mining and Analytics: New Trends is an updated look at the state of technology in the field of data mining and analytics. As processor speeds, databas...
  • Knowledge Discovery Practices and Emerging Applications of Data Mining
    Recent developments have drastically increased the volume and complexity of data available to be mined, leading researchers to explore new ways to glean non-trivial data automatically. Knowledge Discovery Practices and Emerging Applications of Data Mining: Trends and New Domains introduces the reader to recent research activities in the field of data mining. This book covers as...
  • Research and Trends in Data Mining Technologies and Applications
    David Taniar
    ...
  • Developing Metadata Application Profiles
    The prevalence of data science has grown exponentially in recent years. Increases in data exchange have created the need for standards and formats on handling data from different sources. Developing Metadata Application Profiles is an innovative reference source that discusses the latest trends and techniques for effectively managing and exchanging metadata. Including a range o...
  • Modern Technologies for Big Data Classification and Clustering
    Data has increased due to the growing use of web applications and communication devices. It is necessary to develop new techniques of managing data in order to ensure adequate usage. Modern Technologies for Big Data Classification and Clustering is an essential reference source for the latest scholarly research on handling large data sets with conventional data mining and provi...
  • 90 Gelöste Fälle zu Zeitintelligenz in der DAX-Sprache
    Ramón Javier Castro Amador
    Dieser Ratgeber ist rein praktisch ausgerichtet, so dass Sie den gesamten DAX-Code in dieser Publikation anhand einer zum Download verfügbaren .pbix-Datei testen können.'90 gelöste Fälle zu Zeitintelligenz in DAX' ist ein Ratgeber für Benutzer von Microsoft Power BI, der Lösungen für sehr häufige praktische Fälle in Zeitintelligenzmodellen in der Sprache DAX bietet.Um das Verst...
    Disponible

    16,15 €

Otros libros del autor

  • Data Engineering with Apache Spark, Delta Lake, and Lakehouse
    Manoj Kukreja
    Understand the complexities of modern-day data engineering platforms and explore strategies to deal with them with the help of use case scenarios led by an industry expert in big dataKey Features:Become well-versed with the core concepts of Apache Spark and Delta Lake for building data platformsLearn how to ingest, process, and analyze data that can be later used for training m...
    Disponible

    71,99 €