Coming soon, join the waitlist
Data Engineering with Databricks on AWS
Lakehouse engineering with Databricks native tools on AWS.
- Trainer
- To be announced
- Duration
- To be announced
- Mode
- Live online, instructor-led
- Demo timing
- To be announced
About this course
A hands-on workshop on scalable lakehouse architecture with Databricks on AWS. Spark transformations, Delta Lake, governance with Unity Catalog and analytics that bridge data engineering with machine learning.
DatabricksDelta LakeUnity CatalogApache SparkDatabricks SQLMLflow
Who this is for
- Data engineers who know Spark basics and want Databricks skills
- AWS data engineers adding Databricks to their profile
- Analysts and ML engineers who work on shared lakehouse data
What you will be able to do
- Design a medallion lakehouse on Databricks
- Build Delta Lake pipelines with reliable incremental loads
- Govern data access with Unity Catalog
- Schedule and monitor jobs in Databricks
- Serve analytics with Databricks SQL
Syllabus
- 1
Databricks on AWS setup
- Workspace, clusters and compute
- Connecting to S3 securely
- 2
Spark on Databricks
- Notebooks and repos
- Transformations at scale
- 3
Delta Lake
- ACID tables and MERGE
- Optimize, Z-order and vacuum
- 4
Governance
- Unity Catalog
- Lineage and permissions
- 5
Pipelines and jobs
- Declarative pipelines
- Job scheduling and alerts
- 6
Analytics and ML bridge
- Databricks SQL
- MLflow introduction
Projects you will build
Project 1
Medallion pipeline
Bronze, silver and gold layers on Delta Lake with data quality checks at each step.
Questions about this course
When does this batch start?
Dates will be announced soon. Join the waitlist and you will hear first, with an early joiner offer.
Should I finish the AWS course first?
It is not required, but Spark basics help. Our team can guide you on the right order in a free call.
Other courses
Same data. Bigger opportunities.
Book the free live demo or WhatsApp us. We will suggest the right path for your background and goal.
