+
+

Related Products

  • Google Cloud BigQuery
    2,017 Ratings
    Visit Website
  • Google Cloud Platform
    61,012 Ratings
    Visit Website
  • HiveMQ
    91 Ratings
    Visit Website
  • DbVisualizer
    583 Ratings
    Visit Website
  • DataHub
    10 Ratings
    Visit Website
  • Teradata VantageCloud
    1,122 Ratings
    Visit Website
  • AnalyticsCreator
    46 Ratings
    Visit Website
  • StrongDM
    102 Ratings
    Visit Website
  • RaimaDB
    12 Ratings
    Visit Website
  • Couchbase
    412 Ratings
    Visit Website

About

Impala provides low latency and high concurrency for BI/analytic queries on the Hadoop ecosystem, including Iceberg, open data formats, and most cloud storage options. Impala also scales linearly, even in multitenant environments. Impala is integrated with native Hadoop security and Kerberos for authentication, and via the Ranger module, you can ensure that the right users and applications are authorized for the right data. Utilize the same file and data formats and metadata, security, and resource management frameworks as your Hadoop deployment, with no redundant infrastructure or data conversion/duplication. For Apache Hive users, Impala utilizes the same metadata and ODBC driver. Like Hive, Impala supports SQL, so you don't have to worry about reinventing the implementation wheel. With Impala, more users, whether using SQL queries or BI applications, can interact with more data through a single repository and metadata stored from source through analysis.

About

You select the size of the cluster, node capacity, and a set of services, and Yandex Data Proc automatically creates and configures Spark and Hadoop clusters and other components. Collaborate by using Zeppelin notebooks and other web apps via a UI proxy. You get full control of your cluster with root permissions for each VM. Install your own applications and libraries on running clusters without having to restart them. Yandex Data Proc uses instance groups to automatically increase or decrease computing resources of compute subclusters based on CPU usage indicators. Data Proc allows you to create managed Hive clusters, which can reduce the probability of failures and losses caused by metadata unavailability. Save time on building ETL pipelines and pipelines for training and developing models, as well as describing other iterative tasks. The Data Proc operator is already built into Apache Airflow.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Anyone seeking a native analytic database tool for open data and table formats

Audience

Anyone interested in a solution for processing multi-terabyte data arrays

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version
Free Trial

Pricing

$0.19 per hour
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Apache
United States
impala.apache.org

Company Information

Yandex
Founded: 1997
Russia
cloud.yandex.com/en/services/data-proc

Alternatives

Apache Sentry

Apache Sentry

Apache Software Foundation

Alternatives

Apache Iceberg

Apache Iceberg

Apache Software Foundation
Impala

Impala

Command Line Software
CereVoice Me

CereVoice Me

CereProc
esProc

esProc

Raqsoft
SAS Text Miner

SAS Text Miner

SAS Institute

Categories

Categories

Integrations

Apache Hive
Hadoop
3forge
Apache Airflow
Apache Flume
Apache HBase
Apache Iceberg
Apache Spark
Apache Zeppelin
Cloudera Data Warehouse
Data Sentinel
Matplotlib
NumPy
Python
Salesforce Data 360
TensorFlow
Yandex Cloud
Yandex DataSphere
pandas
scikit-image

Integrations

Apache Hive
Hadoop
3forge
Apache Airflow
Apache Flume
Apache HBase
Apache Iceberg
Apache Spark
Apache Zeppelin
Cloudera Data Warehouse
Data Sentinel
Matplotlib
NumPy
Python
Salesforce Data 360
TensorFlow
Yandex Cloud
Yandex DataSphere
pandas
scikit-image
Claim Apache Impala and update features and information
Claim Apache Impala and update features and information
Claim Yandex Data Proc and update features and information
Claim Yandex Data Proc and update features and information