hudi
Upserts, Deletes And Incremental Processing on Big Data.
About this project
Apache Hudi Apache Hudi is an open data lakehouse platform, built on a high-performance open table format to ingest, index, store, serve, transform and manage your data across multiple cloud data environments. Features Hudi stores all data and metadata on cloud storage in open formats, providing the following features across different aspects. Ingestion Built-in ingestion tools for Apache Spark/Apache Flink users. Supports half-dozen file formats, database change logs and streaming data systems. Connect sink for Apache Kafka, to bring external data sources. Storage Optimized storage format, supporting row & columnar data. Timeline metadata to track…
Technologies
Project health
GitHub
Reviews
Built by
Maintain apache/hudi? Claiming verifies admin access through your GitHub account and gives you control of this listing.