All tags
Catalogue tag

#lakehouse

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

8 records
Tagged “lakehouse”Ranked by health index
Go · Maven · PyPI
98Exceptionalhealth index
apache/polaris
Apache Polaris, the interoperable, open source catalog for Apache Iceberg
Java★ 2,011Jul 15, 2026
Apache-2.0Jul 15, 2026 · metrics 2.10.0
Maven · PyPI
97Exceptionalhealth index
prestodb/presto
The official home of the Presto distributed SQL query engine for big data
Java★ 16.7KAug 5, 2026
Apache-2.0Aug 5, 2026 · metrics 2.10.0
Maven · PyPI
96Exceptionalhealth index
StarRocks/starrocks
The world's fastest open query engine for sub-second analytics both on and off the data lakehouse. With the flexibility to support nearly any scenario, StarRocks provides best-in-class performance for multi-dimensional analytics, real-time analytics, and ad-hoc queries. A Linux Foundation project.
Java · C++★ 12KAug 12, 2026
Apache-2.0Aug 12, 2026 · metrics 2.10.0
Go
92Excellenthealth index
datazip-inc/olake
OLake - Fastest Databases, Kafka & S3 Replication to Apache Iceberg with Table optimization (Called OLake Fusion). ⚡ Efficient, quick and scalable data ingestion for real-time analytics. Supported sources : Postgres, MongoDB, MySQL, Oracle, MSSql, DB2, Kafka, S3.
Go★ 1,433Aug 31, 2026
Apache-2.0Aug 31, 2026 · metrics 2.10.0
crates.io · PyPI
89Excellenthealth index
lakekeeper/lakekeeper
Lakekeeper is an Apache-Licensed, secure, fast and easy to use Apache Iceberg REST Catalog written in Rust.
Rust★ 1,399Jul 27, 2026
Apache-2.0Jul 27, 2026 · metrics 2.10.0
npm · PyPI
78Goodhealth index
data-dot-all/dataall
A modern data marketplace that makes collaboration among diverse users (like business, analysts and engineers) easier, increasing efficiency and agility in data projects on AWS.
Python · JavaScript★ 256Aug 16, 2026
Apache-2.0Aug 16, 2026 · metrics 2.10.0
Go · crates.io · Maven +1
65Goodhealth index
ThiagoLange/iceberg-ai-deltalakehouse
AI-Lake Format: self-contained Parquet + HNSW index — unifying tabular data, embeddings, and vector search in a single Iceberg-compatible file
Rust · Python · Kotlin★ 32Jul 18, 2026
Custom licenseJul 18, 2026 · metrics 2.10.0
PyPI
60Moderatehealth index
adidas/lakehouse-engine
The Lakehouse Engine is a configuration driven Spark framework, written in Python, serving as a scalable and distributed engine for several lakehouse algorithms, data flows and utilities for Data Products.
Python★ 294↓ 3,910/moAug 18, 2026
Apache-2.0Aug 18, 2026 · metrics 2.10.0