Microsoft DP-3027: Implement a data engineering solution with Azure Databricks

This course teaches you how to harness the power of Apache Spark and the high-capacity clusters in the platform. Azure Databricks to run complex data engineering tasks in cloudYou will explore streaming processing architectures, implement automated processes, and understand how to optimize performance using Delta Live Tables. In addition, you will learn to orchestrate and monitor data processes through Azure Databricks Jobs, apply governance and security measures to data and integrate Databricks with other services Azure.

Who is it for?

The course is recommended for:

  • Data Engineers who develop large-scale data processing solutions.
  • Data Scientists who need to use Azure Databricks for preparing and processing large data sets.
  • ELT Developers who implement complex data flows in cloud.
  • Professionals who want to learn how to orchestrate, secure, and optimize data processes in Azure Databricks.

What will you learn?

After completing the course, you will know how to:

  • Implement incremental processes using Spark Structured Streaming.
  • Develop streaming architectures with Delta Live Tables.
  • Optimize the performance of data workloads in Spark and Delta Live Tables.
  • Create and manage CI/CD workflows in Azure Databricks.
  • Automate and orchestrate data flows through Azure Databricks Jobs and Azure Data Factory.
  • Manage data security, privacy, and governance with Unity Catalog.
  • You use SQL Warehouses in Azure Databricks for relational queries.
  • run Azure Databricks Notebooks in Azure Data Factory to scale data engineering processes.

Prerequisites:

There are no prerequisites.

Course schedule:

Course materials are in English. Teaching is done in Romanian.

  1. Incremental processing with Spark Structured Streaming
    • Introduction to Spark Structured Streaming
    • Implementing and monitoring incremental processes
  2. Streaming Architectures with Delta Live Tables
    • Architectural models for real-time data
    • Using Delta Live Tables for streaming processes
  3. Optimizing performance with Spark and Delta Live Tables
    • Execution optimization strategies in Spark
    • Increasing the performance of data pipelines
  4. Implementing CI/CD workflows in Azure Databricks
    • Continuous integration and delivery
    • Automate code and process deployment
  5. Automate tasks with Azure Databricks Jobs
    • Creating and scheduling jobs in Azure Databricks
    • Integration with Azure Data Factory and Azure DevOps
    • Monitoring and scaling processes
  6. Data governance and security in Azure Databricks
    • Unity Catalog and data access control
    • Privacy and compliance management
  7. Using SQL Warehouses in Azure Databricks
    • Relational SQL queries on large data sets
    • Optimization of analysis through SQL Warehouses
  8. Running Databricks Notebooks with Azure Data Factory
    • Integrating notebooks into data pipelines
    • Automating data engineering processes at scale cloud

We recommend continuing with:

Certification programs

There are no certification programs at this time.

Microsoft DP-3027: Implement a data engineering solution with Azure Databricks

Personalized offers for groups of at least 2 people

Course details

1
days

Price:

On demand

Delivery:

Classroom Teaching, Hybrid Classroom, Virtual Classroom

Level:

3. Intermediate

Roles:

Data Analyst, Data Engineer, Data Scientist