Try our new research platform with insights from 80,000+ expert users

Amazon EMR vs Cloudera Distribution for Hadoop comparison

 

Comparison Buyer's Guide

Executive Summary

Review summaries and opinions

We asked business professionals to review the solutions they use. Here are some excerpts of what they said:
 

ROI

Sentiment score
6.3
Amazon EMR ROI varies widely, with some companies doubling returns and others reporting up to 20% in cost savings.
Sentiment score
6.1
Assessing ROI from Cloudera for Hadoop varies, with significant value seen in 30 use cases across departments.
 

Customer Service

Sentiment score
7.6
Amazon EMR customer service is generally responsive and knowledgeable, but some experience inconsistencies and slower responses with open-source integrations.
Sentiment score
6.7
Cloudera's customer support is praised for responsiveness and efficiency, despite occasional challenges with less experienced consultants.
The technical support is quite good and better than IBM.
 

Scalability Issues

Sentiment score
7.8
Amazon EMR is scalable and customizable for enterprise applications, despite some performance concerns with resource allocation.
Sentiment score
7.6
Cloudera Distribution for Hadoop is scalable, supporting large deployments and user bases despite challenges in cloud scalability and cost.
 

Stability Issues

Sentiment score
8.1
Amazon EMR is stable and reliable, though occasionally requires configuration adjustments for different data loads.
Sentiment score
7.5
Cloudera Distribution's stability is debated; some experience issues, while others find it reliable, rating it 8-10 out of 10.
 

Room For Improvement

Amazon EMR needs UI improvements, enhanced automation, better stability, debugging, cost management, and advanced features for efficient operations.
Cloudera Distribution for Hadoop struggles with performance, integration, security, and cost, needing enhancements in multiple facets and technologies.
Integrating with Active Directory, managing security, and configuration are the main concerns.
 

Setup Cost

Amazon EMR pricing varies, with potential high costs primarily from infrastructure, despite no standard software licensing fees.
Cloudera Distribution for Hadoop is costly, complex, suitable for large enterprises, and offers a 60-day free trial.
It can be deployed on-premises, unlike competitors' cloud-only solutions.
 

Valuable Features

Amazon EMR offers scalable, efficient data processing with easy integration and management, emphasizing cost-effectiveness, security, and diverse data connectivity.
Cloudera Distribution for Hadoop excels with robust features, security, scalability, AI support, and strong community, surpassing competitors.
This is the only solution that is possible to install on-premise.
 

Categories and Ranking

Amazon EMR
Ranking in Hadoop
3rd
Average Rating
7.8
Reviews Sentiment
7.2
Number of Reviews
22
Ranking in other categories
Cloud Data Warehouse (12th)
Cloudera Distribution for H...
Ranking in Hadoop
2nd
Average Rating
8.0
Reviews Sentiment
6.4
Number of Reviews
50
Ranking in other categories
NoSQL Databases (8th)
 

Mindshare comparison

As of February 2025, in the Hadoop category, the mindshare of Amazon EMR is 13.6%, down from 18.0% compared to the previous year. The mindshare of Cloudera Distribution for Hadoop is 25.7%, up from 22.7% compared to the previous year. It is calculated based on PeerSpot user engagement data.
Hadoop
 

Featured Reviews

Prashant  Singh - PeerSpot reviewer
Seamless data integration enhances reporting efficiency and an easy setup
Amazon EMR has multiple connectors that can connect to various data sources. The service charges are based on processing only, depending on the resources used, which can help save money. It is easy to integrate with other services for storage, allowing data to be shifted to cheaper storage based on usage.
Rok Dolinsek - PeerSpot reviewer
Enables on-premise implementation with powerful data processing capabilities
This is the only solution that is possible to install on-premise. Cloudera provides a hybrid solution that combines compute on cloud or on-premises. It includes all machine learning algorithms in the Spark machine learning library. All functionalities needed for a big data platform and ETL are on the platform, eliminating the need for other tools. It is scalable, ready for vertical scaling, and very powerful, offering numerous functionalities and configurations for generative AI.
report
Use our free recommendation engine to learn which Hadoop solutions are best for your needs.
838,713 professionals have used our research since 2012.
 

Top Industries

By visitors reading reviews
Financial Services Firm
25%
Computer Software Company
13%
Manufacturing Company
9%
Educational Organization
7%
Financial Services Firm
24%
Computer Software Company
14%
Educational Organization
12%
Manufacturing Company
8%
 

Company Size

By reviewers
Large Enterprise
Midsize Enterprise
Small Business
 

Questions from the Community

What do you like most about Amazon EMR?
Amazon EMR is a good solution that can be used to manage big data.
What is your experience regarding pricing and costs for Amazon EMR?
The cost of Amazon EMR is a little bit expensive, especially considering the support package, which includes a gold package.
What needs improvement with Amazon EMR?
Spark jobs take longer on Amazon EMR compared to previous experiences. This aspect could be improved to make them more efficient.
What do you like most about Cloudera Distribution for Hadoop?
The tool can be deployed using different container technologies, which makes it very scalable.
What is your experience regarding pricing and costs for Cloudera Distribution for Hadoop?
The price for Cloudera is average, yet it is very good compared to other solutions. It can be deployed on-premises, unlike competitors' cloud-only solutions.
What needs improvement with Cloudera Distribution for Hadoop?
It is quite complicated to configure and install. Integrating the platform into an information system is always a challenge, especially when starting with on-premise implementation. Integrating wit...
 

Also Known As

Amazon Elastic MapReduce
No data available
 

Overview

 

Sample Customers

Yelp
37signals, Adconion,adgooroo, Aggregate Knowledge, AMD, Apollo Group, Blackberry, Box, BT, CSC
Find out what your peers are saying about Amazon EMR vs. Cloudera Distribution for Hadoop and other solutions. Updated: January 2025.
838,713 professionals have used our research since 2012.