Modern Data Architectures with Python PDF

Modern Data Architectures with Python PDF

Name:
Modern Data Architectures with Python PDF

Published Date:
09/29/2023

Status:
[ Active ]

Description:

Publisher:
PACKT - Packt Publishing, Inc.

Document status:
Active

Format:
Electronic (PDF)

Delivery time:
10 minutes

Delivery time (for Russian version):
200 business days

SKU:

Choose Document Language:
$12
Need Help?
ISBN: 9781801070492

Build scalable and reliable data ecosystems using Data Mesh, Databricks Spark, and Kafka

Key Features:

 * Develop modern data skills used in emerging technologies

 * Learn pragmatic design methodologies such as Data Mesh and data lakehouses

 * Gain a deeper understanding of data governance

 * Purchase of the print or Kindle book includes a free PDF eBook

Book Description:

Modern Data Architectures with Python will teach you how to seamlessly incorporate your machine learning and data science work streams into your open data platforms. You’ll learn how to take your data and create open lakehouses that work with any technology using tried-and-true techniques, including the medallion architecture and Delta Lake.

Starting with the fundamentals, this book will help you build pipelines on Databricks, an open data platform, using SQL and Python. You’ll gain an understanding of notebooks and applications written in Python using standard software engineering tools such as git, pre-commit, Jenkins, and Github. Next, you’ll delve into streaming and batch-based data processing using Apache Spark and Confluent Kafka. As you advance, you’ll learn how to deploy your resources using infrastructure as code and how to automate your workflows and code development. Since any data platform's ability to handle and work with AI and ML is a vital component, you’ll also explore the basics of ML and how to work with modern MLOps tooling. Finally, you’ll get hands-on experience with Apache Spark, one of the key data technologies in today’s market.

By the end of this book, you’ll have amassed a wealth of practical and theoretical knowledge to build, manage, orchestrate, and architect your data ecosystems.

What you will learn:

 * Understand data patterns including delta architecture

 * Discover how to increase performance with Spark internals

 * Find out how to design critical data diagrams

 * Explore MLOps with tools such as AutoML and MLflow

 * Get to grips with building data products in a data mesh

 * Discover data governance and build confidence in your data

 * Introduce data visualizations and dashboards into your data practice

Who this book is for:

This book is for developers, analytics engineers, and managers looking to further develop a data ecosystem within their organization. While they’re not prerequisites, basic knowledge of Python and prior experience with data will help you to read and follow along with the examples.

Authors: Brian Lipp, Michael M. Conti


Edition : 1.
File Size : 1 file , 14 MB
Number of Pages : 318
Published : 09/29/2023
isbn : 9781801070492

History


Related products


Best-Selling Products

CLSI AUTO01-A
Published Date: 12/20/2000
Laboratory Automation: Specimen Container/Specimen Carrier; Approved Standard, AUTO01AE
$54
CLSI AUTO02-A2
Published Date: 01/05/2006
Laboratory Automation: Bar Codes for Specimen Container Identification; Approved Standard, AUTO02A2E
$54
CLSI AUTO03-A2
Published Date: 09/01/2009
Laboratory Automation: Communications with Automated Clinical Laboratory Systems, Instruments, Devices, and Information Systems; Approved Standard, Second Edition, AUTO03A2
$54
CLSI AUTO04-A
Published Date: 03/20/2001
Laboratory Automation: Systems Operational Requirements, Characteristics, and Information Elements; Approved Standard, AUTO04AE
CLSI AUTO05-A
Published Date: 03/20/2001
Laboratory Automation: Electromechanical Interfaces; Approved Standard, AUTO05AE
CLSI AUTO07-A
Published Date: 06/20/2004
Laboratory Automation: Data Content for Specimen Identification; Approved Standard, AUTO07AE