beam

by apache · Data

Apache Beam is a unified programming model for Batch and Streaming data processing.

New0 ratings8,658 starsActive
DataDeveloper ToolJavaPython
Open project ↗↓ Download v2.76.0GitHub⚑ Report
beam preview

About this project

Apache Beam Apache Beam is a unified model for defining both batch and streaming data-parallel processing pipelines, as well as a set of language-specific SDKs for constructing pipelines and Runners for executing them on distributed processing backends, including Apache Flink, Apache Spark, Google Cloud Dataflow, and Hazelcast Jet. 🚀 Quick Start (Beginner Friendly) If you're new to Apache Beam, start here: 1. Choose a language: — Java → Java Quickstart — Python → Python Quickstart — Go → Go Quickstart 2. Run your first example: — Minimal WordCount example (available in this repository) 3. Understand core concepts: — PCollection — PTransform — Pipeline Status …

Technologies

TypeScriptPythonDartJavaGobig-databatchbeam

Project health

Actively maintained
Last updateyesterday
Contributors303
Latest releasev2.76.0
Open issues & PRs3,973
LicenseApache-2.0
On GitHubsince 2016

GitHub

8,658
stars
4,649
forks
303
contributors
3,973
open issues & PRs
Java
language
yesterday
last commit
View on GitHub ↗All releases ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review beam.

Built by

apache
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain apache/beam? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like