Public APIs
Delta Lake favicon

Delta Lake

Cloud Storage & File Sharing

Open-source storage framework enabling Lakehouse architecture with Spark, PrestoDB, Flink, Trino, Hive, and APIs.

Delta Lake's website screenshot

About Delta Lake

Delta Lake is an open-source storage framework that provides ACID transactions and a unified format for data lake tables. It offers a family of libraries for reading and writing Delta tables from different engines and languages, including Delta Spark for Apache Spark, Delta Kernel for building connectors without needing to understand the underlying Delta protocol, Delta Rust for low-level access from Rust (with Python bindings) for use with frameworks like DataFusion and Ballista, and Delta Flink for reading and writing Delta tables from Apache Flink applications. Delta Standalone, a JVM library for reading and writing Delta tables without Spark, is deprecated in favor of Delta Kernel.

The project is aimed at developers building data processing applications or connectors that need to interact with Delta tables, whether through a distributed engine like Spark, Flink, or Trino, or directly from custom applications. API documentation is provided in Scala, Java, and Python depending on the library.

Key features

  • Reads and writes Delta tables using Apache Spark via Delta Spark
  • Delta Kernel provides simple APIs for reading and writing Delta tables without needing to understand the Delta protocol
  • Rust library (with Python bindings) for low-level access to Delta tables
  • Delta Standalone JVM library reads and writes Delta tables without requiring a Spark cluster
  • Delta Flink connector reads and writes data from Apache Flink applications to Delta tables

Frequently asked questions

What languages can I use to work with Delta Lake?

Delta Lake offers APIs for Scala, Java, and Python (via Delta Spark), plus Rust with Python bindings via Delta Rust.

Do I need Apache Spark to use Delta Lake?

No, Delta Standalone and Delta Kernel are JVM libraries that read and write Delta tables without requiring a Spark cluster.

Does Delta Lake integrate with Apache Flink?

Yes, the Delta Flink connector is a JVM library that reads and writes data from Apache Flink applications to Delta tables using the Delta Standalone library.

Is Delta Standalone still recommended?

No, Delta Standalone is deprecated in favor of Delta Kernel, which supports advanced features for reading and writing Delta tables.

Advertise here

Featured products

  • SerpApi - Search API favicon
  • Screenshot Scout favicon
  • TalorData favicon
  • CoreClaw favicon

Show your product to thousands of developers

· 100k monthly pageviews
· 7k newsletter subscribers

Advertise your product