Overview
Data professionals from all disciplines will benefit from this comprehensive introduction to the components of the Databricks Lakehouse Platform that directly support putting ETL pipelines into production. You’ll leverage SQL and Python to define and schedule pipelines that incrementally process new data from a variety of data sources to power analytic applications and dashboards in the lakehouse. This course offers hands-on instruction in Databricks Data Science and Engineering Workspace, Databricks SQL, Delta Live Tables, Databricks Repos, Databricks Task Orchestration and Unity Catalog.
This course will prepare you to take the Databricks Certified Data Engineer Associate exam.
Prerequisites
Participants should have:
- Beginner familiarity with basic cloud concepts (virtual machines, object storage, identity management).
- Ability to perform basic code development tasks (e.g., creating compute instances, running code in notebooks, using basic notebook operations, and importing repositories from Git).
- Intermediate familiarity with SQL, including commands such as CREATE, SELECT, INSERT, UPDATE, DELETE, GROUP BY, JOIN.
- Intermediate experience with SQL concepts such as aggregate functions, filters, sorting, indexes, tables, and views.
- Basic knowledge of Python programming, Jupyter Notebook interface, and PySpark fundamentals.
If you do not have one or more of the pre-requisites QA recommends:
Target Audience
This course is designed for:
- Data Engineers who want to enhance their knowledge of Databricks and Delta Lake.
- Data Analysts looking to expand their expertise in data pipelines and transformation.
- Cloud Engineers and Developers working with big data frameworks.
- Professionals preparing for the Databricks Associate Data Engineering certification.
Delegates will learn how to
- Leverage the Databricks Lakehouse Platform to perform core responsibilities for data pipeline development
- Use SQL and Python to write production data pipelines to extract, transform and load data into tables and views in the lakehouse
- Simplify data ingestion and incremental change propagation using Databricks-native features and syntax, including Delta Live Tables
- Orchestrate production pipelines to deliver fresh results for ad hoc analytics and dashboarding
Outline
Day 1
- Delta Lake
- Relational Entities on Databricks
- ETL with Spark SQL
- Just Enough Python for Spark SQL
- Incremental Data Processing with Structured Streaming and Auto Loader
Day 2
- Medallion Architecture in the Data Lakehouse
- Delta Live Tables
- Task Orchestration with Databricks Jobs
- Databricks SQL
- Managing Permissions in the Lakehouse
- Productionizing Dashboards and Queries on Databricks SQL

Frequently asked questions
How can I create an account on myQA.com?
There are a number of ways to create an account. If you are a self-funder, simply select the "Create account" option on the login page.
If you have been booked onto a course by your company, you will receive a confirmation email. From this email, select "Sign into myQA" and you will be taken to the "Create account" page. Complete all of the details and select "Create account".
If you have the booking number you can also go here and select the "I have a booking number" option. Enter the booking reference and your surname. If the details match, you will be taken to the "Create account" page from where you can enter your details and confirm your account.
Find more answers to frequently asked questions in our FAQs: Bookings & Cancellations page.
How do QA’s virtual classroom courses work?
Our virtual classroom courses allow you to access award-winning classroom training, without leaving your home or office. Our learning professionals are specially trained on how to interact with remote attendees and our remote labs ensure all participants can take part in hands-on exercises wherever they are.
We use the WebEx video conferencing platform by Cisco. Before you book, check that you meet the WebEx system requirements and run a test meeting to ensure the software is compatible with your firewall settings. If it doesn’t work, try adjusting your settings or contact your IT department about permitting the website.
How do QA’s online courses work?
QA online courses, also commonly known as distance learning courses or elearning courses, take the form of interactive software designed for individual learning, but you will also have access to full support from our subject-matter experts for the duration of your course. When you book a QA online learning course you will receive immediate access to it through our e-learning platform and you can start to learn straight away, from any compatible device. Access to the online learning platform is valid for one year from the booking date.
All courses are built around case studies and presented in an engaging format, which includes storytelling elements, video, audio and humour. Every case study is supported by sample documents and a collection of Knowledge Nuggets that provide more in-depth detail on the wider processes.
When will I receive my joining instructions?
Joining instructions for QA courses are sent two weeks prior to the course start date, or immediately if the booking is confirmed within this timeframe. For course bookings made via QA but delivered by a third-party supplier, joining instructions are sent to attendees prior to the training course, but timescales vary depending on each supplier’s terms. Read more FAQs.
When will I receive my certificate?
Certificates of Achievement are issued at the end the course, either as a hard copy or via email. Read more here.