Try our new research platform with insights from 80,000+ expert users

Cloudera Distribution for Hadoop vs HPE Ezmeral Data Fabric comparison

 

Comparison Buyer's Guide

Executive Summary

Review summaries and opinions

We asked business professionals to review the solutions they use. Here are some excerpts of what they said:
 

Categories and Ranking

Cloudera Distribution for H...
Ranking in Hadoop
2nd
Average Rating
8.0
Reviews Sentiment
6.4
Number of Reviews
49
Ranking in other categories
NoSQL Databases (8th)
HPE Ezmeral Data Fabric
Ranking in Hadoop
5th
Average Rating
8.0
Reviews Sentiment
6.1
Number of Reviews
12
Ranking in other categories
No ranking in other categories
 

Mindshare comparison

As of February 2025, in the Hadoop category, the mindshare of Cloudera Distribution for Hadoop is 25.7%, up from 22.7% compared to the previous year. The mindshare of HPE Ezmeral Data Fabric is 14.5%, up from 10.7% compared to the previous year. It is calculated based on PeerSpot user engagement data.
Hadoop
 

Featured Reviews

Miodrag-Stanic - PeerSpot reviewer
You can manage all services from one place in an integrated manner
We switched to Airflow because Cloudera is outdated. It's not widely used. It would be good if we had the Spark 3.5. Spark is quite old. Cloudera is now offering an alternate solution as a replacement for AWS. AWS works badly with small files. The solution is not fit for on-premise distributions. It should be containerized so we can deploy it as containers within Kubernetes. We had one upgrade from CDH to CDP, which lasted for a long time. And I would expect with containerized deployment, it would be upgraded much more quickly than we had the experience.
Arnab Chatterjee - PeerSpot reviewer
It's flexible and easily accessible across multiple locations, but the upgrade process is complicated
Upgrading Ezmeral to a new version is a pain. They're trying to make the solution more container-friendly, so I think they're going in the right direction. The only problem we've had in the past was the upgrades. The process isn't smooth due to how the Red Hat operating system upgrades currently work. They're transforming their host stack to increase cloud readiness and edge compute capability. HPE is transitioning from a standard data-driven approach to one powered by AI analytics. That's something they have released very recently. I haven't tried that, but it will probably make things easier. The ability to adapt Ezmeral to the public cloud is probably missing. I've heard that they're getting leaner. However, it doesn't have a clear managed services offering for you if you want to deploy this stack on the cloud. That's a problem. This probably won't meet your needs if you require consistency across on-prem and the cloud. It's not Ezmeral's fault. None of the products would fit the bill. Cloud offerings are biased towards their own implementation. It's a general issue on most big data platforms. They're already working towards that, but it hasn't been released.

Quotes from Members

We asked business professionals to review the solutions they use. Here are some excerpts of what they said:
 

Pros

"Very good end-to-end security features."
"The data science aspect of the solution is valuable."
"The file system is a valuable feature."
"We experienced many issues when we started working with Hadoop 3.0 in the Cloudera 6.0 version, so there are a lot of things that need to improve. I believe they are working on that."
"Cloudera, as a whole, is designed to provide organizations with solutions for big data."
"CDH has a wide variety of proprietary tools that we use, like Impala. So from that perspective, it's quite useful as opposed to something open-source. We get a lot of value from Cloudera's proprietary tools."
"The solution is reliable and stable, it fits our requirements."
"The features I find most valuable is that the solution is that it is easy to install and to work with. It starts with the installation and from there on the management is very simple and centralized."
"I like the administration part."
"HPE Ezmeral Data Fabric can be accessed from any namespace globally as you would access it from a machine using an NFS."
"My customers find the product cheaper compared to other solutions. The previous solution that we used did not have unified analytics like the runtime or the analog."
"It is a stable solution...It is a scalable solution."
"The model creation was very interesting, especially with the libraries provided by the platform."
 

Cons

"There are multiple bugs when we update."
"Currently, we are using many other tools such as Spark and Blade Job to improve the performance."
"Without the big data environment, we cannot store all of this data live. We have billions of records and terabytes of storage to be used. It's not an option actually for us to have a big data environment."
"The dashboard could be improved."
"The competitors provide better functionalities."
"The solution does not support multiple languages very well and this means users need to create work-arounds to implement some solutions."
"The pricing needs to improve."
"The performance of some analytics engines provided by Cloudera is not that good."
"The product is not user-friendly."
"The deployment could be faster. I want more support for the data lake in the next release."
"HPE Ezmeral Data Fabric is not compatible with third-party tools."
"Upgrading Ezmeral to a new version is a pain. They're trying to make the solution more container-friendly, so I think they're going in the right direction. The only problem we've had in the past was the upgrades. The process isn't smooth due to how the Red Hat operating system upgrades currently work."
"Having the ability to extend the services provided by the platform to an API architecture, a micro-services architecture, could be very helpful."
 

Pricing and Cost Advice

"I haven't bought a license for this solution. I'm only using the Apache license version."
"Cloudera requires a license to use."
"The price is very high. The solution is expensive."
"Cloudera Distribution for Hadoop is expensive, with support costs involved."
"I believe we pay for a three-year license."
"When comparing with Oracle Sybase and SQL, it's cheaper. It's not expensive."
"I wouldn't recommend CDH to others because of its high cost."
"The product’s price depends from project to project."
"There is a need for my company to pay for the licensing costs of the solution."
"The tool's price is cheap and based on a usage basis. The solution's licensing costs are yearly and there are no extra costs."
"HPE is flexible with you if you are an existing customer. They offer different models that might be beneficial for your organization. It all depends on how you negotiate."
report
Use our free recommendation engine to learn which Hadoop solutions are best for your needs.
832,138 professionals have used our research since 2012.
 

Top Industries

By visitors reading reviews
Financial Services Firm
23%
Computer Software Company
14%
Educational Organization
11%
Manufacturing Company
9%
Financial Services Firm
18%
Computer Software Company
14%
Government
7%
Retailer
7%
 

Company Size

By reviewers
Large Enterprise
Midsize Enterprise
Small Business
 

Questions from the Community

What do you like most about Cloudera Distribution for Hadoop?
The tool can be deployed using different container technologies, which makes it very scalable.
What is your experience regarding pricing and costs for Cloudera Distribution for Hadoop?
The tool is expensive. Overall, it's not a cheap software tool, and that is why only large enterprises who are mature enough and have an architecture that is complex enough opt for Cloudera, as its...
What needs improvement with Cloudera Distribution for Hadoop?
The tool doesn't support reporting, and relational databases are still the major source of reporting data. Apache Iceberg will be launched soon within the Cloudera cluster for analytical purposes. ...
What do you like most about HPE Ezmeral Data Fabric?
It is a stable solution...It is a scalable solution.
What needs improvement with HPE Ezmeral Data Fabric?
There are some drawbacks in HPE Ezmeral Data Fabric when it comes to the interoperability part. HPE Ezmeral Data Fabric is not compatible with third-party tools. For example, HPE Ezmeral Data Fabri...
What is your primary use case for HPE Ezmeral Data Fabric?
The main purpose of HPE Ezmeral Data Fabric for me is that it acts as a database. In my company, we store our data with the help of HPE Ezmeral Data Fabric. It is possible to use Spark engine with ...
 

Also Known As

No data available
MapR, MapR Data Platform
 

Overview

 

Sample Customers

37signals, Adconion,adgooroo, Aggregate Knowledge, AMD, Apollo Group, Blackberry, Box, BT, CSC
Valence Health, Goodgame Studios, Pico, Terbium Labs, sovrn, Harte Hanks, Quantium, Razorsight, Novartis, Experian, Dentsu ix, Pontis Transitions, DataSong, Return Path, RAPP, HP
Find out what your peers are saying about Cloudera Distribution for Hadoop vs. HPE Ezmeral Data Fabric and other solutions. Updated: January 2025.
832,138 professionals have used our research since 2012.