Apache SparkApache Software Foundation
|
R2 SQLCloudflare
|
|||||
Related Products
|
||||||
About
Apache Spark™ is a unified analytics engine for large-scale data processing. Apache Spark achieves high performance for both batch and streaming data, using a state-of-the-art DAG scheduler, a query optimizer, and a physical execution engine. Spark offers over 80 high-level operators that make it easy to build parallel apps. And you can use it interactively from the Scala, Python, R, and SQL shells. Spark powers a stack of libraries including SQL and DataFrames, MLlib for machine learning, GraphX, and Spark Streaming. You can combine these libraries seamlessly in the same application. Spark runs on Hadoop, Apache Mesos, Kubernetes, standalone, or in the cloud. It can access diverse data sources. You can run Spark using its standalone cluster mode, on EC2, on Hadoop YARN, on Mesos, or on Kubernetes. Access data in HDFS, Alluxio, Apache Cassandra, Apache HBase, Apache Hive, and hundreds of other data sources.
|
About
R2 SQL is Cloudflare’s serverless, distributed analytics query engine (currently in open beta) that enables you to run SQL queries over Apache Iceberg tables stored in R2 Data Catalog without needing to manage your own compute clusters. It is built to efficiently query large volumes of data by leveraging metadata pruning, partition-level statistics, file and row-group filtering, and Cloudflare’s globally distributed compute infrastructure to parallelize execution. The system works by integrating with R2 object storage and an Iceberg catalog layer, so you can ingest data via Cloudflare Pipelines into Iceberg tables, and then query that data with minimal overhead. Queries can be issued via the Wrangler CLI or HTTP API (with an API token granting permissions across R2 SQL, Data Catalog, and storage). During the open beta period, using R2 SQL itself is not billed, only storage and standard R2 operations incur charges.
|
|||||
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
|||||
Audience
Organizations that want a unified analytics engine for large-scale data processing
|
Audience
Data engineers and backend developers needing a solution to run scalable analytics over large datasets stored in R2 using SQL
|
|||||
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Not Supported
|
Support
Phone Support
Supported
24/7 Live Support
Not Supported
Online
Supported
|
|||||
API
Offers API
Not Supported
|
API
Offers API
Supported
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
No information available.
Free Version
Supported
Free Trial
Not Supported
|
Pricing
Free
Free Version
Supported
Free Trial
Not Supported
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Not Supported
|
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Supported
|
|||||
Company InformationApache Software Foundation
Founded: 1999
United States
spark.apache.org
|
Company InformationCloudflare
Founded: 2009
United States
developers.cloudflare.com/r2-sql/
|
|||||
Alternatives |
Alternatives |
|||||
|
|
||||||
|
|
|
|||||
|
|
||||||
Categories |
Categories |
|||||
Streaming Analytics Features
Data Enrichment
Supported
Data Wrangling / Data Prep
Supported
Multiple Data Source Support
Supported
Process Automation
Supported
Real-time Analysis / Reporting
Not Supported
Visualization Dashboards
Not Supported
|
||||||
Integrations
Apache Iceberg
Supported
SQL
Supported
Apache Kylin
Supported
Azure HDInsight
Supported
Gable
Supported
Gemini Enterprise Agent Platform
Supported
Google Cloud Bigtable
Supported
HStreamDB
Supported
IBM Analytics for Apache Spark
Supported
MLflow
Supported
|
Integrations
Apache Iceberg
Supported
SQL
Supported
Apache Kylin
Not Supported
Azure HDInsight
Not Supported
Gable
Not Supported
Gemini Enterprise Agent Platform
Not Supported
Google Cloud Bigtable
Not Supported
HStreamDB
Not Supported
IBM Analytics for Apache Spark
Not Supported
MLflow
Not Supported
|
|||||
|
|
|