Building ETL Pipelines with SQL - Mandarin Chinese
本课程讲授如何在 Databricks Data Intelligence Platform 上使用纯 SQL 构建生产就绪的 ETL 管道。学员将学习以下内容:使用 Auto Loader 的流式处理表(用于增量摄取)、具有增量 刷新 功能的物化视图(用于白银到黄金的转换)、AUTO CDC(FLOW AUTO CDC)(用于声明式 SCD Type 1 和 Type 2 维度管理),以及 Apache Spark Declarative Pipelines(SDP)——将上述对象组合成具有数据质量期望、PRIVATE 表和统一监控功能的生产多表管道。本课程以一个真实的零售数据集为例,贯穿金银铜架构(铜 → 白银 → 黄金)的完整流程。使用 Lakeflow Jobs 进行编排作为附加模块提供。
注意:对于 SCORM 课程文件,请在完成内容后关闭 SCORM 窗口。请勿点击“下一课”按钮,否则可能导致 SCORM 模块无法标记为已完成。
本课程内容面向具备以下技能、知识和能力的学员:
• 导航 Databricks 工作区(侧边栏、Catalog Explorer、SQL Editor)
• Unity Catalog 基础知识(目录、模式、表、卷)
• 中级 SQL(SELECT、JOIN、GROUP BY、CAST、COALESCE、CREATE TABLE)
• 数据仓库概念(事实表/维度表、星型模式、金银铜架构)
• 对 ETL 工作流的基本了解
Self-Paced
Custom-fit learning paths for data, analytics, and AI roles and career paths through on-demand videos
Registration options
Databricks has a delivery method for wherever you are on your learning journey
Self-Paced
Custom-fit learning paths for data, analytics, and AI roles and career paths through on-demand videos
Register nowInstructor-Led
Public and private courses taught by expert instructors across half-day to two-day courses
Register nowBlended Learning
Self-paced and weekly instructor-led sessions for every style of learner to optimize course completion and knowledge retention. Go to Subscriptions Catalog tab to purchase
Purchase nowSkills@Scale
Comprehensive training offering for large scale customers that includes learning elements for every style of learning. Inquire with your account executive for details

