Implement data engineering solutions using Azure Databricks (DP-750T00)
Build, secure, govern, deploy, and optimize scalable lakehouse solutions using Azure Databricks, Unity Catalog, Lakeflow, and Apache Spark for data engineering.
Register or Request Training
- Private class for your team
- Live expert instructor
- Online or on‑location
- Customizable agenda
- Proposal responses same day as request
Course Overview
Master end-to-end data engineering with Azure Databricks and Unity Catalog. Learn to configure environments, build robust ingestion and transformation pipelines, implement enterprise governance and security, and deploy optimized workloads. By the end of this course, you will be prepared to implement, secure, monitor, and maintain scalable lakehouse solutions.
Course Benefits
- Configure Azure Databricks compute, storage, and integrations for data engineering workloads.
- Create, organize, secure, and govern data assets with Unity Catalog.
- Design data models, partitioning schemes, clustering strategies, and slowly changing dimensions.
- Ingest batch and streaming data using Lakeflow, notebooks, SQL, Auto Loader, and Spark Structured Streaming.
- Cleanse, transform, and load data while managing schema drift and data quality constraints.
- Design, build, schedule, and manage Azure Databricks pipelines and Lakeflow Jobs.
- Apply Git-based development, testing, packaging, and deployment practices.
- Monitor, troubleshoot, and optimize Azure Databricks workloads.
Delivery Methods
Live expert-led online training from anywhere. Guaranteed to run .
Delivered for your team at your site or online.
Course Outline
- Explore Azure Databricks
- Get started with Azure Databricks
- Identify Azure Databricks workloads
- Understand key concepts
- Explore data governance using Unity Catalog and Microsoft Purview
- Module assessment
- Understand Azure Databricks Architecture
- Understand Azure Databricks architecture
- Understand Unity Catalog managed storage
- Understand external storage
- Understand default storage
- Module assessment
- Understand Azure Databricks Integrations
- Understand integration with Microsoft Fabric
- Understand integration with Power BI
- Understand integration with VS Code
- Understand integration with Power Platform
- Understand integration with Copilot Studio
- Understand integration with Microsoft Purview
- Understand integration with Microsoft Foundry
- Module assessment
- Select and Configure Compute in Azure Databricks
- Choose an appropriate compute type
- Configure compute performance
- Configure compute features
- Install libraries for compute
- Configure compute access
- Module assessment
- Create and Organize Objects in Unity Catalog
- Apply naming conventions
- Create a catalog
- Create a schema
- Create tables and views
- Create volumes
- Implement DDL operations
- Implement a foreign catalog
- Configure AI/BI Genie instructions
- Secure Unity Catalog Objects
- Understand the query lifecycle
- Implement access control strategies
- Understand fine-grained access control
- Implement row filtering and column masking
- Access Azure Key Vault secrets
- Authenticate data access with service principals
- Authenticate resource access with managed identities
- Module assessment
- Govern Unity Catalog Objects
- Create and preserve table definitions
- Configure attribute-based access control with tags and policies
- Apply data retention policies
- Set up and manage data lineage
- Configure audit logging
- Design a secure Delta Sharing strategy
- Module assessment
- Design and Implement Data Modeling with Azure Databricks
- Design ingestion logic and data source configuration
- Choose a data ingestion tool
- Choose a data table format
- Design and implement a data partitioning scheme
- Choose a slowly changing dimension type
- Implement a type 2 slowly changing dimension
- Design and implement a temporal table to record changes over time
- Choose column or table granularity based on requirements
- Choose between managed and external tables
- Design and implement a clustering strategy
- Ingest Data into Unity Catalog
- Ingest data with Lakeflow Connect
- Ingest data with notebooks
- Ingest data with SQL methods
- Ingest data with a change data capture feed
- Ingest data with Spark Structured Streaming
- Ingest data with Auto Loader
- Ingest data with Lakeflow Spark Declarative Pipelines
- Module assessment
- Cleanse, Transform, and Load Data into Unity Catalog
- Profile data
- Choose column data types
- Resolve duplicates and nulls
- Transform data with filters and aggregations
- Transform data with joins and set operators
- Transform data with denormalization and pivots
- Load data with merge, insert, and append operations
- Module assessment
- Implement and Manage Data Quality Constraints with Azure Databricks
- Implement validation checks
- Implement data type checks
- Detect and manage schema drift
- Manage data quality with pipeline expectations
- Module assessment
- Design and Implement Data Pipelines with Azure Databricks
- Design the order of operations for a pipeline
- Choose between notebooks and Lakeflow Pipelines
- Design Lakeflow job logic
- Design error handling for pipelines and jobs
- Create a pipeline with a notebook
- Create a pipeline with Lakeflow Spark Declarative Pipelines
- Module assessment
- Implement Lakeflow Jobs with Azure Databricks
- Create job setup and configuration
- Configure job triggers
- Schedule a job
- Configure job alerts
- Configure automatic restarts
- Module assessment
- Implement Development Lifecycle Processes in Azure Databricks
- Apply Git version control best practices
- Manage branching and pull requests
- Implement a testing strategy
- Configure and package Declarative Automation Bundles
- Deploy a bundle with the Databricks CLI
- Module assessment
- Monitor, Troubleshoot, and Optimize Workloads in Azure Databricks
- Monitor and manage cluster consumption
- Troubleshoot and repair Lakeflow Jobs
- Troubleshoot Spark jobs and notebooks
- Investigate caching, skewing, spilling, and shuffling
- Implement log streaming with Azure Log Analytics
- Module assessment
Class Materials
Each student receives a comprehensive set of materials, including course notes and all class examples.
Class Prerequisites
Experience in the following is required for this Azure class:
- Ability to work with SQL.
- Experience using Python and notebooks for data engineering tasks.
- Understanding of Azure Databricks workspaces and Unity Catalog.
- Familiarity with data access patterns and core data engineering and data warehouse concepts.
Experience in the following would be useful for this Azure class:
- Fundamental knowledge of data analytics concepts.
- Basic understanding of cloud storage and data organization principles.
- Foundational knowledge of Azure security, including Microsoft Entra ID.
- Familiarity with Git version control fundamentals.
Have questions about this course?
We can help with curriculum details, delivery options, pricing, or anything else. Reach out and we’ll point you in the right direction.