Skip to main content

Get Started with Databricks Platform Administration

In this course, you will learn the basics of platform administration on the Databricks Data Intelligence Platform. It offers a comprehensive overview of the Unity Catalog, a vital component for effective data governance within Databricks environments. Divided into five modules, it begins with a detailed introduction to Databricks infrastructure and its data intelligence platform, including an in-depth walkthrough of the Databricks Workspace. You will explore data governance principles within Unity Catalog, covering its key concepts, architecture, and roles. The course further emphasizes managing Unity Catalog metastores and compute resources, including clusters and SQL warehouses. Finally, you'll master data access control by learning about privileges, fine-grained access, and how to govern data objects. By the end, you will be equipped with essential skills to administer the Unity Catalog to implement effective data governance, optimize compute resources, and enforce robust data security strategies.



Languages Available: English | 日本語 | Português BR | 한국어


Skill Level
Onboarding
Duration
2h
Prerequisites

The content was developed for participants with these skills/knowledge/abilities:

• Familiarity with the Databricks Data Intelligence Platform and basic workspace operations (create clusters, run code in notebooks, use basic notebook operations)

• Basic understanding of identity and access management concepts (users, groups, service principals, authentication, authorization)

• Understanding of Unity Catalog fundamentals including the hierarchical object model (metastore, catalogs, schemas, tables, volumes, models)

• Basic knowledge of data governance principles and access control concepts (permissions, entitlements, administrative responsibilities)

• Beginner familiarity with cloud computing concepts (virtual machines, object storage, identity management, cloud resources)

• Basic understanding of workspace administration concepts including user management and permission assignment

• Knowledge of account-level versus workspace-level administration and the relationship between them

• A basic understanding of cloud computing and SQL concepts, including networking, SQL queries, and database structures such as tables and views

• Familiarity with Python programming, the Jupyter notebook interface, and foundational PySpark operations

Outline

1. Databricks Overview

• Databricks Data Intelligence Platform

• Demo: Databricks Workspace Walkthrough


2. Databricks Platform Administration

2.1 Data Governance in Unity Catalog

• Data Governance Overview

• Unity Catalog Key Concepts

• Databricks Roles

• Databricks Identities


2.2 Managing Principals in Unity Catalog

• Managing Principals in Unity Catalog - Overview

• Demo: Adding and Deleting Users

• Demo: Adding and Deleting Service Principals

• Demo: Adding and Deleting Groups

• Demo: Assigning Users, Service Principals, and Groups to Workspaces


2.3 Managing Unity Catalog Metastores

• Demo: Creating and Deleting Metastores in Unity Catalog

• Demo: Assigning a Metastore to a Workspace in Unity Catalog

• Demo: Assigning Metastore Administrators in Unity Catalog


2.4 Compute Resources and Unity Catalog

• Clusters

• Demo: Creating a Cluster in Unity Catalog

• SQL Warehouses

• Demo: Creating SQL Warehouses in Unity Catalog


2.5 Data Access Control in Unity Catalog

• Privileges in Unity Catalog

• Fine-grained Access Control

• Demo: Implementing Fine-Grained Access Control in Unity Catalog

Upcoming Public Classes

Date
Time
Your Local Time
Language
Price
Aug 06
03 PM - 05 PM (Europe/London)
-
English
Free
Aug 12
09 AM - 11 AM (America/Los_Angeles)
-
English
Free
Aug 19
09 AM - 11 AM (America/Los_Angeles)
-
English
Free
Sep 09
03 PM - 05 PM (Europe/London)
-
English
Free
Sep 18
09 AM - 11 AM (America/Los_Angeles)
-
English
Free
Oct 14
03 PM - 05 PM (Europe/London)
-
English
Free
Oct 22
09 AM - 11 AM (America/Los_Angeles)
-
English
Free

Public Class Registration

If your company has purchased success credits or has a learning subscription, please fill out the Training Request form. Otherwise, you can register below.

Private Class Request

If your company is interested in private training, please submit a request.

See all our registration options

Registration options

Databricks has a delivery method for wherever you are on your learning journey

Runtime

Self-Paced

Custom-fit learning paths for data, analytics, and AI roles and career paths through on-demand videos

Register now

Instructors

Instructor-Led

Public and private courses taught by expert instructors across half-day to two-day courses

Register now

Learning

Blended Learning

Self-paced and weekly instructor-led sessions for every style of learner to optimize course completion and knowledge retention. Go to Subscriptions Catalog tab to purchase

Purchase now

Scale

Skills@Scale

Comprehensive training offering for large scale customers that includes learning elements for every style of learning. Inquire with your account executive for details

Upcoming Public Classes

Data Engineer

Data Ingestion with Lakeflow Connect

This course provides a comprehensive introduction to Lakeflow Connect, a scalable and simplified solution for ingesting data into Databricks from a wide range of sources. You’ll begin by exploring the different types of Lakeflow Connect connectors (Standard and Managed) and learn various data ingestion techniques, including batch, incremental batch, and streaming ingestion. You'll also review the key benefits of using Delta table and the Medallion architecture

Next, you’ll develop practical skills for ingesting data from cloud object storage using Lakeflow Connect Standard Connectors. This includes working with methods such as CREATE TABLE AS SELECT (CTAS), COPY INTO, and Auto Loader, with an emphasis on the benefits and considerations of each approach. You’ll also learn how to append metadata columns to your bronze-level tables during ingestion into the Databricks Data Intelligence Platform. The course then covers how to handle records that don’t match your table schema using the rescued data column, along with strategies for managing and analyzing this data. You’ll also explore techniques for ingesting and flattening semi-structured JSON data.

Following this, you’ll explore how to perform enterprise-grade data ingestion using Lakeflow Connect Managed Connectors to bring in data from databases and Software-as-a-Service (SaaS) applications. The course also introduces Partner Connect as an option for integrating partner tools into your ingestion workloads.

Finally, the course wraps up with alternative ingestion strategies, including MERGE INTO operations and leveraging the Databricks Marketplace, equipping you with a strong foundation to support modern data engineering use cases.

Note: For SCORM lecture files, please ensure that you close the SCORM window after completing the content. Do not click the ‘Next Lesson’ button, as doing so may prevent the SCORM module from being marked as complete.

Free
2h
Associate

Questions?

If you have any questions, please refer to our Frequently Asked Questions page.