Aller au contenu principal
SynapseLink
Data Engineering — Foundations
Data Engineering

Data Engineering — Foundations

Lay the foundations of data engineering: SQL, modeling and your first pipelines.

This course introduces the building blocks of data engineering: SQL, modeling, Python and ETL pipeline concepts. As a capstone, you build a mini-ETL that extracts, transforms and loads a dataset.

Objectives

  • Write SQL queries to manipulate data
  • Model data according to analytical needs
  • Automate simple processing in Python
  • Understand how an ETL pipeline works

Program

SQL and modeling

  • Advanced queries and joins
  • Normalization and keys
  • Models suited to analysis

Python for data

  • File-processing scripts
  • Reading and writing sources
  • Automating repetitive tasks

ETL pipeline concepts

  • Extract, transform, load
  • Orchestrating simple steps
  • Building an end-to-end mini-ETL

Prerequisites

Programming and database basics appreciated.