Generative AI Engineering with Databricks

This course is aimed at data scientists, machine learning engineers, and other data practitioners who want to build generative AI applications using the latest and most popular frameworks and Databricks capabilities.

Note: Databricks Academy is transitioning to a notebook-based format for classroom sessions within the Databricks environment, discontinuing the use of slide decks for lectures in the first two modules. You can access the lecture notebooks in the Vocareum lab environment.

Below, we describe each of the four, four-hour modules included in this course.

Building Retrieval Agents On Databricks: This course provides hands-on training for building retrieval agents on the Databricks Data Intelligence Platform. Participants will learn to parse unstructured documents into structured data, transform and chunk content for retrieval workflows, build vector search solutions for document retrieval, and develop production-ready agents using MLflow and Agent Bricks. The course covers the complete agent lifecycle from document processing through embedding generation, vector indexing, and agent deployment with governance capabilities.

Building Single-Agent Applications on Databricks: This course provides hands-on training for building single-agent applications on the Databricks Data Intelligence Platform. Students will learn to create AI agents that leverage Unity Catalog functions as tools, implement comprehensive tracing and monitoring with MLflow, and deploy agents using both traditional frameworks like LangChain and modern solutions like Agent Bricks. The course covers the complete agent lifecycle from initial tool creation and testing in AI Playground through production deployment with governance, evaluation, and continuous improvement capabilities.

Generative AI Application Evaluation and Governance: This is your introduction to evaluating and governing generative AI systems. First, you’ll explore the meaning behind and motivation for building evaluation and governance/security systems. Next, we’ll connect evaluation and governance systems to the Databricks Data Intelligence Platform. Third, we’ll teach you about a variety of evaluation techniques for specific components and types of applications. Finally, the course will conclude with an analysis of evaluating entire AI systems with respect to performance and cost.

Generative AI Application Deployment and Monitoring: Ready to learn how to deploy, operationalize, and monitor generative deploying, operationalizing, and monitoring generative AI applications? This module will help you gain skills in the deployment of generative AI applications using tools like Model Serving. We’ll also cover how to operationalize generative AI applications following best practices and recommended architectures. Finally, we’ll discuss the idea of monitoring generative AI applications and their components using Lakehouse Monitoring.

Skill Level

Associate

Duration

16h

Prerequisites

• Ability to write production-quality Python code, including OOP, exception handling, decorators, type hints, and proper documentation.

• Experience writing advanced SQL SELECT queries, handling data types and NULL values, and creating reusable, well-documented SQL functions.

• Comfort navigating the Databricks workspace and notebooks, managing compute, using Catalog Explorer, and understanding Databricks-managed services.

• Understanding of LLM behavior, basic prompt engineering, RAG concepts, agent reasoning, and working with REST APIs and JSON payloads.

• Basic familiarity with MLflow, agent frameworks (e.g., LangChain), and recommended Databricks training such as AI Agents Fundamentals.

• Familiarity with natural language processing concepts

• Familiarity with prompt engineering/prompt engineering best practices

• Familiarity with the Databricks Data Intelligence Platform

• Familiarity with RAG (preparing data, building a RAG architecture, concepts like embedding, vectors, vector databases, etc.)

• Experience with building LLM applications using multi-stage reasoning LLM chains and agents

• Experience with Databricks Data Intelligence Platform tools for evaluation and governance.

• Understanding of Unity Catalog concepts including catalogs and schemas

• Basic knowledge of MLflow

Outline

Building Retrieval Agents On Databricks

• Document Parsing and Chunking

• Vector Search for Retrieval

• Building and Logging Retrieval Agents

• Agent Bricks

Building Single-Agent Applications on Databricks

• Foundations of Agents

• Building Single Agents

• Reproducible Agents

• Production-Ready Agents with Agent Bricks

Generative AI Application Evaluation and Governance

• Importance of Evaluating GenAI Applications

• Securing and Governing GenAI Applications

• GenAI Evaluation Techniques

• End-to-end Application Evaluation

Generative AI Application Deployment and Monitoring

• Model Deployment Fundamentals

• Batch Deployment

• Real-Time Deployment

• AI System Monitoring

• LLMOps Concepts

Upcoming Public Classes

Date	Time	Language	Price
Date	Time	Language	Price	Mar 17 - 20	11 AM - 03 PM (Asia/Singapore)	English	$1500.00
Mar 18 - 19	09 AM - 05 PM (America/New_York)	English	$1500.00
Mar 19 - 20	09 AM - 05 PM (Europe/London)	English	$1500.00
Apr 14 - 15	09 AM - 05 PM (America/New_York)	English	$1500.00
Apr 16 - 17	09 AM - 05 PM (Europe/London)	English	$1500.00
Apr 21 - 24	11 AM - 03 PM (Asia/Singapore)	English	$1500.00

Public Class Registration

If your company has purchased success credits or has a learning subscription, please fill out the Training Request form. Otherwise, you can register below.

Customer registration Partner registration

Private Class Request

If your company is interested in private training, please submit a request.

Request Private Training

See all our registration options

Registration options

Databricks has a delivery method for wherever you are on your learning journey

Self-Paced

Custom-fit learning paths for data, analytics, and AI roles and career paths through on-demand videos

Instructor-Led

Public and private courses taught by expert instructors across half-day to two-day courses

Blended Learning

Self-paced and weekly instructor-led sessions for every style of learner to optimize course completion and knowledge retention. Go to Subscriptions Catalog tab to purchase

Purchase now

Skills@Scale

Comprehensive training offering for large scale customers that includes learning elements for every style of learning. Inquire with your account executive for details

Upcoming Public Classes

Data Governance at Scale

In this course, you will learn how to implement data governance at scale on Databricks using Unity Catalog, with a focus on attribute-based access control, observability, and federated sharing. You will configure ABAC with governed tags, migrate from legacy fine-grained controls, enable and use system tables for audit and cost monitoring, deploy Lakehouse Monitoring for data and model quality, interpret lineage for impact and compliance, and apply federated governance and Delta Sharing patterns for secure cross-cloud collaboration.

Note: Databricks Academy is transitioning to a notebook-based format for classroom sessions within the Databricks environment, discontinuing the use of slide decks for lectures. You can access the lecture notebooks in the Vocareum lab environment.

SQL Analytics on Databricks

In this course, you'll learn how to effectively use Databricks for data analytics, with a specific focus on Databricks SQL. As a Databricks Data Analyst, your responsibilities will include finding relevant data, analyzing it for potential applications, and transforming it into formats that provide valuable business insights.

You will also understand your role in managing data objects and how to manipulate them within the Databricks Data Intelligence Platform, using tools such as Notebooks, the SQL Editor, and Databricks SQL.

Additionally, you will learn about the importance of Unity Catalog in managing data assets and the overall platform. Finally, the course will provide an overview of how Databricks facilitates performance optimization and teach you how to access Query Insights to understand the processes occurring behind the scenes when executing SQL analytics on Databricks.

Languages Available: English | 日本語 | Português BR | 한국어

Apache Spark Developer

Introduction to Apache Spark™

This course offers essential knowledge of Apache Spark, with a focus on its distributed architecture and practical applications for large-scale data processing. Participants will explore programming frameworks, learn the Spark DataFrame API, and develop skills for reading, writing, and transforming data using Python-based Spark workflows.

Languages Available: English | 日本語 | 한국어