Databricks Lakehouse Platform Cookbook
eBook - ePub

Databricks Lakehouse Platform Cookbook

100+ recipes for building a scalable and secure Databricks Lakehouse (English Edition)

Dr. Alan L. Dennis

Share book
  1. English
  2. ePUB (mobile friendly)
  3. Available on iOS & Android
eBook - ePub

Databricks Lakehouse Platform Cookbook

100+ recipes for building a scalable and secure Databricks Lakehouse (English Edition)

Dr. Alan L. Dennis

Book details
Table of contents
Citations

About This Book

Analyze, Architect, and Innovate with Databricks Lakehouse

Key Features
? Create a Lakehouse using Databricks, including ingestion from source to Bronze.
? Refinement of Bronze items to business-ready Silver items using incremental methods.
? Construct Gold items to service the needs of various business requirements.

Description
The Databricks Lakehouse is groundbreaking technology that simplifies data storage, processing, and analysis. This cookbook offers a clear and practical guide to building and optimizing your Lakehouse to make data-driven decisions and drive impactful results.This definitive guide walks you through the entire Lakehouse journey, from setting up your environment, and connecting to storage, to creating Delta tables, building data models, and ingesting and transforming data. We start off by discussing how to ingest data to Bronze, then refine it to produce Silver. Next, we discuss how to create Gold tables and various data modeling techniques often performed in the Gold layer. You will learn how to leverage Spark SQL and PySpark for efficient data manipulation, apply Delta Live Tables for real-time data processing, and implement Machine Learning and Data Science workflows with MLflow, Feature Store, and AutoML. The book also delves into advanced topics like graph analysis, data governance, and visualization, equipping you with the necessary knowledge to solve complex data challenges.By the end of this cookbook, you will be a confident Lakehouse expert, capable of designing, building, and managing robust data-driven solutions.

What you will learn
? Design and build a robust Databricks Lakehouse environment.
? Create and manage Delta tables with advanced transformations.
? Analyze and transform data using SQL and Python.
? Build and deploy machine learning models for actionable insights.
? Implement best practices for data governance and security.

Who this book is for
This book is meant for Data Engineers, Data Analysts, Data Scientists, Business intelligence professionals, and Architects who want to go to the next level of Data Engineering using the Databricks platform to construct Lakehouses.

Table of Contents
1. Introduction to Databricks Lakehouse
2. Setting-up a Databricks Workspace
3. Connecting to Storage
4. Creating Delta Tables
5. Data Profiling and Modeling in the Lakehouse
6. Extracting from Source and Loading to Bronze
7. Transforming to Create Silver
8. Transforming to Create Gold for Business Purposes
9. Machine Learning and Data Science
10. SQL Analysis
11. Graph Analysis
12. Visualizations
13. Governance
14. Operations
15. Tips, Tricks, Troubleshooting, and Best Practices

Frequently asked questions

How do I cancel my subscription?
Simply head over to the account section in settings and click on “Cancel Subscription” - it’s as simple as that. After you cancel, your membership will stay active for the remainder of the time you’ve paid for. Learn more here.
Can/how do I download books?
At the moment all of our mobile-responsive ePub books are available to download via the app. Most of our PDFs are also available to download and we're working on making the final remaining ones downloadable now. Learn more here.
What is the difference between the pricing plans?
Both plans give you full access to the library and all of Perlego’s features. The only differences are the price and subscription period: With the annual plan you’ll save around 30% compared to 12 months on the monthly plan.
What is Perlego?
We are an online textbook subscription service, where you can get access to an entire online library for less than the price of a single book per month. With over 1 million books across 1000+ topics, we’ve got you covered! Learn more here.
Do you support text-to-speech?
Look out for the read-aloud symbol on your next book to see if you can listen to it. The read-aloud tool reads text aloud for you, highlighting the text as it is being read. You can pause it, speed it up and slow it down. Learn more here.
Is Databricks Lakehouse Platform Cookbook an online PDF/ePUB?
Yes, you can access Databricks Lakehouse Platform Cookbook by Dr. Alan L. Dennis in PDF and/or ePUB format, as well as other popular books in Informatik & Programmierung in Python. We have over one million books available in our catalogue for you to explore.

Information

Year
2023
ISBN
9789355519566

Table of contents