spark
Apache Spark - A unified analytics engine for large-scale data processing
About this project
Apache Spark Spark is a unified analytics engine for large-scale data processing. It provides high-level APIs in Scala, Java, Python, and R (Deprecated), and an optimized engine that supports general computation graphs for data analysis. It also supports a rich set of higher-level tools including Spark SQL for SQL and DataFrames, pandas API on Spark for pandas workloads, MLlib for machine learning, GraphX for graph processing, and Structured Streaming for stream processing. — Official version: — Development version: Online Documentation You can find the latest Spark documentation, including a programming guide, on the project web page. This README file only contains basic…
Technologies
Project health
GitHub
Reviews
Built by
Maintain apache/spark? Claiming verifies admin access through your GitHub account and gives you control of this listing.