全部标签
目录标签

#lakehouse

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

8 条记录
标签为“lakehouse”按健康指数排序
Go · Maven · PyPI
98卓越健康指数
apache/polaris
Apache Polaris, the interoperable, open source catalog for Apache Iceberg
Java★ 2,0112026年7月15日
Apache-2.02026年7月15日 · 指标 2.10.0
Maven · PyPI
97卓越健康指数
prestodb/presto
The official home of the Presto distributed SQL query engine for big data
Java★ 16.7K2026年8月5日
Apache-2.02026年8月5日 · 指标 2.10.0
Maven · PyPI
96卓越健康指数
StarRocks/starrocks
The world's fastest open query engine for sub-second analytics both on and off the data lakehouse. With the flexibility to support nearly any scenario, StarRocks provides best-in-class performance for multi-dimensional analytics, real-time analytics, and ad-hoc queries. A Linux Foundation project.
Java · C++★ 12K2026年8月12日
Apache-2.02026年8月12日 · 指标 2.10.0
Go
92优秀健康指数
datazip-inc/olake
OLake - Fastest Databases, Kafka & S3 Replication to Apache Iceberg with Table optimization (Called OLake Fusion). ⚡ Efficient, quick and scalable data ingestion for real-time analytics. Supported sources : Postgres, MongoDB, MySQL, Oracle, MSSql, DB2, Kafka, S3.
Go★ 1,4332026年8月31日
Apache-2.02026年8月31日 · 指标 2.10.0
crates.io · PyPI
89优秀健康指数
lakekeeper/lakekeeper
Lakekeeper is an Apache-Licensed, secure, fast and easy to use Apache Iceberg REST Catalog written in Rust.
Rust★ 1,3992026年7月27日
Apache-2.02026年7月27日 · 指标 2.10.0
npm · PyPI
78良好健康指数
data-dot-all/dataall
A modern data marketplace that makes collaboration among diverse users (like business, analysts and engineers) easier, increasing efficiency and agility in data projects on AWS.
Python · JavaScript★ 2562026年8月16日
Apache-2.02026年8月16日 · 指标 2.10.0
Go · crates.io · Maven +1
65良好健康指数
ThiagoLange/iceberg-ai-deltalakehouse
AI-Lake Format: self-contained Parquet + HNSW index — unifying tabular data, embeddings, and vector search in a single Iceberg-compatible file
Rust · Python · Kotlin★ 322026年7月18日
自定义许可证2026年7月18日 · 指标 2.10.0
PyPI
60中等健康指数
adidas/lakehouse-engine
The Lakehouse Engine is a configuration driven Spark framework, written in Python, serving as a scalable and distributed engine for several lakehouse algorithms, data flows and utilities for Data Products.
Python★ 294↓ 3,910/月2026年8月18日
Apache-2.02026年8月18日 · 指标 2.10.0