
Data Engineering
Data Engineering — Expert
Master big data with Spark, Databricks and streaming at scale.
You process large volumes with Spark and Databricks, set up streaming and structure a governed data lake. The course prepares for the Azure DP-203 certification via a complete data platform.
Objectives
- Process large volumes with Spark and Databricks
- Set up streaming processing
- Structure a reliable and governed data lake
- Prepare for the Azure DP-203 certification
Program
Spark and Databricks
- Distributed processing and optimization
- Databricks notebooks and jobs
- Delta Lake and table reliability
Streaming and data lake
- Real-time ingestion
- Layered data lake architecture
- Partitioning and columnar formats
Quality, governance and DP-203
- Quality controls and data testing
- Catalog, lineage and security
- Review and DP-203 mock exam
Prerequisites
Experience with data pipelines and Python.