Try our new research platform with insights from 80,000+ expert users

Cloudera Distribution for Hadoop vs HPE Ezmeral Data Fabric comparison

 

Comparison Buyer's Guide

Executive Summary

Review summaries and opinions

We asked business professionals to review the solutions they use. Here are some excerpts of what they said:
 

Categories and Ranking

Cloudera Distribution for H...
Ranking in Hadoop
2nd
Average Rating
8.0
Reviews Sentiment
6.4
Number of Reviews
50
Ranking in other categories
NoSQL Databases (8th)
HPE Ezmeral Data Fabric
Ranking in Hadoop
4th
Average Rating
8.0
Reviews Sentiment
6.1
Number of Reviews
12
Ranking in other categories
No ranking in other categories
 

Mindshare comparison

As of April 2025, in the Hadoop category, the mindshare of Cloudera Distribution for Hadoop is 25.0%, up from 23.0% compared to the previous year. The mindshare of HPE Ezmeral Data Fabric is 15.0%, up from 10.6% compared to the previous year. It is calculated based on PeerSpot user engagement data.
Hadoop
 

Featured Reviews

Rok Dolinsek - PeerSpot reviewer
Enables on-premise implementation with powerful data processing capabilities
This is the only solution that is possible to install on-premise. Cloudera provides a hybrid solution that combines compute on cloud or on-premises. It includes all machine learning algorithms in the Spark machine learning library. All functionalities needed for a big data platform and ETL are on the platform, eliminating the need for other tools. It is scalable, ready for vertical scaling, and very powerful, offering numerous functionalities and configurations for generative AI.
Arnab Chatterjee - PeerSpot reviewer
It's flexible and easily accessible across multiple locations, but the upgrade process is complicated
Upgrading Ezmeral to a new version is a pain. They're trying to make the solution more container-friendly, so I think they're going in the right direction. The only problem we've had in the past was the upgrades. The process isn't smooth due to how the Red Hat operating system upgrades currently work. They're transforming their host stack to increase cloud readiness and edge compute capability. HPE is transitioning from a standard data-driven approach to one powered by AI analytics. That's something they have released very recently. I haven't tried that, but it will probably make things easier. The ability to adapt Ezmeral to the public cloud is probably missing. I've heard that they're getting leaner. However, it doesn't have a clear managed services offering for you if you want to deploy this stack on the cloud. That's a problem. This probably won't meet your needs if you require consistency across on-prem and the cloud. It's not Ezmeral's fault. None of the products would fit the bill. Cloud offerings are biased towards their own implementation. It's a general issue on most big data platforms. They're already working towards that, but it hasn't been released.

Quotes from Members

We asked business professionals to review the solutions they use. Here are some excerpts of what they said:
 

Pros

"The most valuable feature is Impala, the querying engine, which is very fast."
"We had a data warehouse before all the data. We can process a lot more data structures."
"Cloudera provides a hybrid solution that combines compute on cloud or on-premises."
"Provides a viable open-source solution for enterprise implementations and reliable, intelligent data analysis."
"Cloudera is a very manageable solution with good support."
"The product provides better data processing features than other tools."
"The tool's most interesting features are the distributed file system and unstructured data processing capability. Because we have a lot of unstructured data, like XML and social media logs, these features make it more valuable than the usual data warehousing solutions."
"We experienced many issues when we started working with Hadoop 3.0 in the Cloudera 6.0 version, so there are a lot of things that need to improve. I believe they are working on that."
"My customers find the product cheaper compared to other solutions. The previous solution that we used did not have unified analytics like the runtime or the analog."
"It is a stable solution...It is a scalable solution."
"HPE Ezmeral Data Fabric can be accessed from any namespace globally as you would access it from a machine using an NFS."
"I like the administration part."
"The model creation was very interesting, especially with the libraries provided by the platform."
 

Cons

"The user infrastructure and user interface needs to be improved, as well as the performance. The GUI needs to be better."
"The one thing that we struggled with predominately was support. Because it was relatively new, support was always a big issue and I think it's still a bit of an ongoing concern with the team currently managing it."
"Without the big data environment, we cannot store all of this data live. We have billions of records and terabytes of storage to be used. It's not an option actually for us to have a big data environment."
"It is quite complicated to configure and install."
"This is a very expensive solution."
"Cloudera Distribution for Hadoop has a limited feature list and a lot of costs involved."
"The dashboard could be improved."
"While the deployed product is generally functional, there are instances where it presents difficulties."
"The deployment could be faster. I want more support for the data lake in the next release."
"Having the ability to extend the services provided by the platform to an API architecture, a micro-services architecture, could be very helpful."
"HPE Ezmeral Data Fabric is not compatible with third-party tools."
"The product is not user-friendly."
"Upgrading Ezmeral to a new version is a pain. They're trying to make the solution more container-friendly, so I think they're going in the right direction. The only problem we've had in the past was the upgrades. The process isn't smooth due to how the Red Hat operating system upgrades currently work."
 

Pricing and Cost Advice

"Cloudera requires a license to use."
"I believe we pay for a three-year license."
"I wouldn't recommend CDH to others because of its high cost."
"When comparing with Oracle Sybase and SQL, it's cheaper. It's not expensive."
"The tool is not expensive."
"The tool is expensive...For the SMB market or customers whose environments are not that complex and do not have multiple systems running, Cloudera might not be a good option."
"It is an expensive product."
"The pricing must be improved."
"There is a need for my company to pay for the licensing costs of the solution."
"The tool's price is cheap and based on a usage basis. The solution's licensing costs are yearly and there are no extra costs."
"HPE is flexible with you if you are an existing customer. They offer different models that might be beneficial for your organization. It all depends on how you negotiate."
report
Use our free recommendation engine to learn which Hadoop solutions are best for your needs.
845,040 professionals have used our research since 2012.
 

Top Industries

By visitors reading reviews
Financial Services Firm
25%
Computer Software Company
15%
Educational Organization
12%
Manufacturing Company
7%
Financial Services Firm
19%
Computer Software Company
16%
Retailer
8%
Comms Service Provider
7%
 

Company Size

By reviewers
Large Enterprise
Midsize Enterprise
Small Business
 

Questions from the Community

What do you like most about Cloudera Distribution for Hadoop?
The tool can be deployed using different container technologies, which makes it very scalable.
What is your experience regarding pricing and costs for Cloudera Distribution for Hadoop?
The price for Cloudera is average, yet it is very good compared to other solutions. It can be deployed on-premises, unlike competitors' cloud-only solutions.
What needs improvement with Cloudera Distribution for Hadoop?
It is quite complicated to configure and install. Integrating the platform into an information system is always a challenge, especially when starting with on-premise implementation. Integrating wit...
What do you like most about HPE Ezmeral Data Fabric?
It is a stable solution...It is a scalable solution.
What needs improvement with HPE Ezmeral Data Fabric?
There are some drawbacks in HPE Ezmeral Data Fabric when it comes to the interoperability part. HPE Ezmeral Data Fabric is not compatible with third-party tools. For example, HPE Ezmeral Data Fabri...
What is your primary use case for HPE Ezmeral Data Fabric?
The main purpose of HPE Ezmeral Data Fabric for me is that it acts as a database. In my company, we store our data with the help of HPE Ezmeral Data Fabric. It is possible to use Spark engine with ...
 

Also Known As

No data available
MapR, MapR Data Platform
 

Overview

 

Sample Customers

37signals, Adconion,adgooroo, Aggregate Knowledge, AMD, Apollo Group, Blackberry, Box, BT, CSC
Valence Health, Goodgame Studios, Pico, Terbium Labs, sovrn, Harte Hanks, Quantium, Razorsight, Novartis, Experian, Dentsu ix, Pontis Transitions, DataSong, Return Path, RAPP, HP
Find out what your peers are saying about Cloudera Distribution for Hadoop vs. HPE Ezmeral Data Fabric and other solutions. Updated: March 2025.
845,040 professionals have used our research since 2012.