Overview
This is an introductory course that serves as an appropriate entry point to learn Data Engineering with Databricks.
Prerequisites
- Beginner familiarity with basic cloud concepts (virtual machines, object storage, identity management)
- Ability to perform basic code development tasks (create compute, run code in notebooks, use basic notebook operations, import repos from git, etc.)
- Intermediate familiarity with basic SQL concepts (CREATE, SELECT, INSERT, UPDATE, DELETE, WHILE, GROUP BY, JOIN, etc.)
- Intermediate experience with basic SQL concepts such as SQL commands, aggregate functions, filters and sorting, indexes, tables, and views.
- Basic knowledge of Python programming, jupyter notebook interface, and PySpark fundamentals.
Delegates will learn how to
Data Ingestion with Delta Lake
- This course is designed for Data Engineers to deepen their understanding of Delta Lake to handle data ingestion, transformation, and management with ease. Using the latest features of Delta Lake, learners will explore real-world applications to enhance data workflows, optimize performance, and ensure data reliability.
Deploy Workloads with Databricks Workflows
- This course is designed for data engineer professionals who are looking to leverage Databricks for streamlined and efficient data workflows. By the end of this course, you’ll be well-versed in using Databricks' Jobs and Workflows functionalities to automate, manage, and monitor complex data pipelines. The course includes hands-on labs and best practices to ensure a deep understanding and practical ability to manage workflows in production environments.
Build Data Pipelines with Delta Live Tables
- This comprehensive course is designed to understand the Medallion Architecture using Delta Live Tables. Participants will learn how to create robust and efficient data pipelines for structured and unstructured data, understand the nuances of managing data quality, and unlock the potential of Delta Live Tables. By the end of this course, participants will have hands-on experience building pipelines, troubleshooting issues, and monitoring their data flows within the Delta Live Tables environment.
Data Management and Governance with Unity Catalog
- In this course, you'll learn about data management and governance using Databricks Unity Catalog. It covers foundational concepts of data governance, complexities in managing data lakes, Unity Catalog's architecture, security, administration, and advanced topics like fine-grained access control, data segregation, and privilege management.
*This course seeks to prepare students to complete the Associate Data Engineering certification exam, and provides the requisite knowledge to take the course Advanced Data Engineering with Databricks.
Outline
Module 1: Data Ingestion with Delta Lake
- Delta Lake and Data Objects
- Set Up and Load Delta Tables
- Basic Transformations
- Load Data Lab
- Cleaning Data
- Complex Transformations
- SQL UDFs
- Advanced Delta Lake Features
- Manipulate Delta Tables Lab
Module 2: Deploy Workloads with Databricks Workflows
- Introduction to Workflows
- Jobs Compute
- Scheduling Tasks with the Jobs UI
- Workflows Lab
- Jobs Features
- Explore Scheduling Options
- Conditional Tasks and Repairing Runs
- Modular Orchestration
- Databricks Workflows Best Practices
Module 3: Build Data Pipelines with Delta Live Tables
- The Medallion Architecture
- Introduction to Delta Live Tables
- Using the Delta Live Tables UI
- SQL Pipelines
- Python Pipelines
- Delta Live Tables Running Modes
- Pipeline Results
- Pipeline Event Logs
- Optional - Land New Data
Module 4: Data Management and Governance with Unity Catalog
- Data Governance Overview
- Demo: Populating the Metastore
- Lab: Navigating the Metastore
- Organization and Access Patterns
- Demo: Upgrading Tables to Unity Catalog
- Security and Administration in Unity Catalog
- Databricks Marketplace Overview
- Privileges in Unity Catalog
- Demo: Controlling Access to Data
- Fine-Grained Access Control
- Lab: Migrating and Managing Data in Unity Catalog
QA reserves the right to improve the specification and format of its courses for the benefit of its customers without notice to the customer.
Frequently asked questions
How can I create an account on myQA.com?
There are a number of ways to create an account. If you are a self-funder, simply select the "Create account" option on the login page.
If you have been booked onto a course by your company, you will receive a confirmation email. From this email, select "Sign into myQA" and you will be taken to the "Create account" page. Complete all of the details and select "Create account".
If you have the booking number you can also go here and select the "I have a booking number" option. Enter the booking reference and your surname. If the details match, you will be taken to the "Create account" page from where you can enter your details and confirm your account.
Find more answers to frequently asked questions in our FAQs: Bookings & Cancellations page.
How do QA’s virtual classroom courses work?
Our virtual classroom courses allow you to access award-winning classroom training, without leaving your home or office. Our learning professionals are specially trained on how to interact with remote attendees and our remote labs ensure all participants can take part in hands-on exercises wherever they are.
We use the WebEx video conferencing platform by Cisco. Before you book, check that you meet the WebEx system requirements and run a test meeting to ensure the software is compatible with your firewall settings. If it doesn’t work, try adjusting your settings or contact your IT department about permitting the website.
How do QA’s online courses work?
QA online courses, also commonly known as distance learning courses or elearning courses, take the form of interactive software designed for individual learning, but you will also have access to full support from our subject-matter experts for the duration of your course. When you book a QA online learning course you will receive immediate access to it through our e-learning platform and you can start to learn straight away, from any compatible device. Access to the online learning platform is valid for one year from the booking date.
All courses are built around case studies and presented in an engaging format, which includes storytelling elements, video, audio and humour. Every case study is supported by sample documents and a collection of Knowledge Nuggets that provide more in-depth detail on the wider processes.
When will I receive my joining instructions?
Joining instructions for QA courses are sent two weeks prior to the course start date, or immediately if the booking is confirmed within this timeframe. For course bookings made via QA but delivered by a third-party supplier, joining instructions are sent to attendees prior to the training course, but timescales vary depending on each supplier’s terms. Read more FAQs.
When will I receive my certificate?
Certificates of Achievement are issued at the end the course, either as a hard copy or via email. Read more here.