Amazon EMR

Amazon EMR

Amazon
+
+

Related Products

  • Teradata VantageCloud
    1,122 Ratings
    Visit Website
  • JS7 JobScheduler
    1 Rating
    Visit Website
  • HiveMQ
    91 Ratings
    Visit Website
  • Apify
    1,441 Ratings
    Visit Website
  • Google Cloud Platform
    61,011 Ratings
    Visit Website
  • Juspay
    17 Ratings
    Visit Website
  • Source Defense
    7 Ratings
    Visit Website
  • wp2print
    7,598 Ratings
    Visit Website
  • KrakenD
    71 Ratings
    Visit Website
  • Parasoft
    148 Ratings
    Visit Website

About

Amazon EMR is the industry-leading cloud big data platform for processing vast amounts of data using open-source tools such as Apache Spark, Apache Hive, Apache HBase, Apache Flink, Apache Hudi, and Presto. With EMR you can run Petabyte-scale analysis at less than half of the cost of traditional on-premises solutions and over 3x faster than standard Apache Spark. For short-running jobs, you can spin up and spin down clusters and pay per second for the instances used. For long-running workloads, you can create highly available clusters that automatically scale to meet demand. If you have existing on-premises deployments of open-source tools such as Apache Spark and Apache Hive, you can also run EMR clusters on AWS Outposts. Analyze data using open-source ML frameworks such as Apache Spark MLlib, TensorFlow, and Apache MXNet. Connect to Amazon SageMaker Studio for large-scale model training, analysis, and reporting.

About

The Stackable data platform was designed with openness and flexibility in mind. It provides you with a curated selection of the best open source data apps like Apache Kafka, OpenSearch, Trino, and Apache Spark. While other current offerings either push their proprietary solutions or deepen vendor lock-in, Stackable takes a different approach. All data apps work together seamlessly and can be added or removed in no time. Based on Kubernetes, it runs everywhere, on-prem or in the cloud. stackablectl and a Kubernetes cluster are all you need to run your first stackable data platform. Within minutes, you will be ready to start working with your data. Configure your one-line startup command right here. Similar to kubectl, stackablectl is designed to easily interface with the Stackable Data Platform. Use the command line utility to deploy and manage stackable data apps on Kubernetes. With stackablectl, you can create, delete, and update components.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Companies that want to easily run and scale Apache Spark, Hive, Presto, and other big data frameworks

Audience

Enterprises wanting a solution to deploy and run their data platforms on their sovereign Kubernetes.

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

No information available.
Free Version
Free Trial

Pricing

Free
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Amazon
Founded: 1994
United States
aws.amazon.com/emr/

Company Information

Stackable
Founded: 2020
Germany
stackable.tech/

Alternatives

Alternatives

Canvas Credentials

Canvas Credentials

Instructure
E-MapReduce

E-MapReduce

Alibaba
Hercules

Hercules

Leisure Holding
Apache Spark

Apache Spark

Apache Software Foundation

Categories

Categories

Data Management Features

Customer Data
Data Analysis
Data Capture
Data Integration
Data Migration
Data Quality Control
Data Security
Information Governance
Master Data Management
Match & Merge

Data Warehouse Features

Ad hoc Query
Analytics
Data Integration
Data Migration
Data Quality Control
ETL - Extract / Transfer / Load
In-Memory Processing
Match & Merge

Integrations

Apache HBase
Apache Hive
Apache Spark
Amazon SageMaker Data Wrangler
Apache Kafka
Apache NiFi
Feast
Gurucul
IBM watsonx.data integration
Lyftrondata
MinIO
Pelanor
Protegrity
SAS Studio
Service Center
Sifflet
Tonic Ephemeral
Trino
Zepl

Integrations

Apache HBase
Apache Hive
Apache Spark
Amazon SageMaker Data Wrangler
Apache Kafka
Apache NiFi
Feast
Gurucul
IBM watsonx.data integration
Lyftrondata
MinIO
Pelanor
Protegrity
SAS Studio
Service Center
Sifflet
Tonic Ephemeral
Trino
Zepl
Claim Amazon EMR and update features and information
Claim Amazon EMR and update features and information
Claim Stackable and update features and information
Claim Stackable and update features and information