by Matt Jones, Paul Lappas and Richard Tomlinson
Today we are announcing the general availability of Delta Live Tables (DLT) on Google Cloud. DLT pipelines empower data engineers to build reliable, streaming and batch data pipelines effortlessly across their preferred cloud environments.
DLT provides a simplified framework for creating production-ready data pipelines to incrementally ingest, transform, and process data at scale.
Running data pipelines in production involves a lot of operational code to take care of activities like task orchestration, checkpointing, automatic restart on failure, performance optimization, data exception handling, and so on. This code is very complex to develop and time-consuming to maintain, leaving many data practitioners with limited bandwidth to focus on their core mission of transforming data to deliver new insights.
DLT addresses these challenges by providing a simple, declarative approach to data pipeline development while automating the complex operational components associated with running these pipelines reliably in production. DLT simplifies ETL development and operations, allowing analysts, data scientists, and data engineers to focus on delivering value from data.

BioIntelliSense has been one of our preview customers for Delta Live Tables on Google Cloud, with over 100 DLT pipelines in production. The BioIntelliSense Data-as-a-Service (DaaS) platform, and its continuous health monitoring and clinical intelligence solutions for in-hospital to home, rely on high-quality, tightly-governed data.
Dnyanesh Sonavane, Director of Data Engineering, summarizes the impact that DLT on Google Cloud has had for his organization:
Our Data Engineering team is focused on passive data collection through medical-grade wearable technology that results in high resolution data for our customers and our internal data science and analytics team. As we scale our business, data sources and data sets grow exponentially, making it more complicated to implement and maintain data transformation and lineage.
DLT has sped up our ELT / ETL pipeline development with declarative definitions reducing time for data delivery. The DAGs inferred by DLT, along with automated infrastructure management, empower data traceability making troubleshooting easy. —Dnyanesh Sonavane, Director of Data Engineering, BioIntelliSense
DLT pipelines provide two powerful, but easy to use, primitives for data processing: streaming tables (STs) and materialized views (MVs).
Streaming Tables are the ideal way to bring data into the "bronze" layer of your medallion architecture. With a single SQL statement or a few lines Python of code, users can scalably ingest data from various sources such as cloud storage (S3, ADLS, GCS), message buses (EventHub, Kafka, Kinesis), and more. This ingestion occurs incrementally, enabling low-latency and cost-effective pipelines, without the need for managing complex infrastructure.

Materialized Views are ideal for transforming data in the "silver" and "gold" layers of your medallion architecture. With MVs, users can simply define a query and any aggregations or transformation, and MVs incrementally refresh to achieve low latencies.

DLT pipelines are a powerful automation framework that make it easy to define a series of STs and MVs and chain them together to form data processing pipelines. With DLT pipelines, a user can simply define the queries for STs and MVs, and DLT pipelines enable real-time understanding of dependencies and automate tasks such as orchestration, retries, recovery, auto-scaling, and performance optimization. This allows data engineers to treat their data as code and leverage modern software engineering best practices like testing, error-handling, monitoring, and documentation to deploy reliable pipelines at scale.
DLT pipelines also offer powerful features and APIs to manage data quality and advanced data modeling capabilities like change-data-capture and slowly-changing dimensions (SCD).
Building and running DLT pipelines on Google Cloud has many advantages. Here are just a few:
By expanding the availability of DLT pipelines to Google Cloud, Databricks reinforces its commitment to providing a partner-friendly ecosystem while offering customers the flexibility to choose the cloud platform that best suits their needs. Whether you're on AWS, Azure, or Google Cloud, DLT empowers data engineers to accelerate ETL development, ensure high data quality, and unify batch and streaming workloads.
To learn more about Delta Live Tables, visit here or start your free trial today.
Delta Live Tables (DLT) is a simplified, declarative framework for building production-ready data pipelines that incrementally ingest, transform, and process data at scale. It automates the complex operational work — such as task orchestration, checkpointing, automatic restart on failure, and performance optimization — that data engineers would otherwise have to build and maintain by hand. This lets data practitioners spend more time transforming data into insights rather than writing and maintaining operational code.
Streaming tables (STs) and materialized views (MVs) are the two core primitives DLT uses for data processing. Streaming tables incrementally ingest data into the bronze layer of a medallion architecture from sources like cloud storage, EventHub, Kafka, and Kinesis, while materialized views incrementally refresh transformations and aggregations in the silver and gold layers. Most DLT pipelines combine STs for ingestion with MVs for transformation.
DLT pipelines can be built using familiar Python or SQL code, and as of this announcement they run on Google Cloud in addition to AWS and Azure. This gives data engineers the flexibility to develop pipelines using the language they already know while choosing the cloud platform that best fits their organization. DLT also includes a broad ecosystem of streaming connectors, including support for Google Pub/Sub.
DLT automates operational tasks such as orchestration, retries, recovery, autoscaling, and performance optimization for pipelines built from streaming tables and materialized views. Automated error handling and restart capabilities help ensure pipeline resilience and minimize downtime. DLT also supports comprehensive testing and CI/CD capabilities so teams can apply software engineering best practices like testing, error-handling, monitoring, and documentation to their pipelines.
DLT offers built-in APIs for change-data-capture and slowly-changing dimensions (SCD), along with features to manage data quality throughout a pipeline. Its data quality monitoring proactively identifies and addresses issues affecting data integrity as pipelines run. These capabilities are available across all DLT primitives, whether pipelines are built with streaming tables or materialized views.
Subscribe to our blog and get the latest posts delivered to your inbox.