Skip to main content
LAKEHOUSE STORAGE

Built for open, intelligent data storage

Choose your storage location and format, with full ownership and portability of your data.
Lakehouse Storage
TOP TEAMS SUCCEED WITH DATA INTELLIGENCE
benefits

Flexible, fast lakehouse storage with open formats

Eliminate data management headaches with open table formats, centralized governance and automatic data optimizations.

Delta Lake and Apache Iceberg™ format compatibility

A single copy of source data in Delta Lake or Apache Iceberg™ that can be accessed by any engine.

Unified governance across data and AI assets

A single catalog for data discovery and governance, across your data and AI assets.

AI-driven optimization for speed and low cost

AI-powered models autonomously optimize and maintain data for speed and low cost.

Features

Choose your storage location and open table format

Choose the storage location and open format that works for you. Keep your data portable, without vendor lock-in.

Best-in-class read and write performance for Delta Lake and Apache Iceberg™ tables, out of the box, with storage optimizations not available in any other lakehouse.

Managed Tables

Additional lakehouse storage features

ACID Transactions

Atomicity, consistency, isolation and durability guarantees provided by open table format protocols.

Predictive Optimization

AI-driven table optimizations based on your data and usage patterns that keep your tables tuned, automatically.

Liquid Clustering

Out-of-the-box, self-tuning data layout that scales with your data — no partitions required.

Change Data Feed

Track row-level changes between versions of a Delta table.

Time Travel

Historical information about tables lets you audit operations, roll back a table or query a table at a specific point in time.

Structured Streaming

Integration with Apache Spark™ Structured Streaming, a near real-time processing engine that offers end-to-end fault tolerance with exactly-once processing guarantees.

USE CASES

Lakehouse storage for ETL, warehousing and sharing

ETL Process

Build and manage reliable data pipelines

Managed tables act as both batch tables and a streaming source and sink. Streaming data ingest, batch historic backfill and interactive queries all work out of the box and directly integrate with Spark Structured Streaming.

related products

Discover, govern and share your data and AI assets

Learn more about how the Databricks Data + AI Platform empowers your data teams across all your data and AI workloads.

Unity Catalog

The industry’s only unified and open governance solution for data and AI, built into the Databricks Data + AI Platform.

Delta Sharing

The first open source approach to data sharing across data, analytics and AI. Securely share live data across platforms, clouds and regions.

Data + AI Platform

Explore the full range of tools available on the Databricks Data + AI Platform to seamlessly integrate data and AI across your organization.

Learn more

Lakehouse storage FAQ

Open table formats are open source, standard table formats that enable ACID transactions on data lakes. Databricks supports Delta Lake and Apache Iceberg™. This means a single copy of your data can be written once and read by any compatible engine, rather than being locked into a proprietary format.

Ready to become a data + AI company?

Take the first steps in your transformation