Try our new research platform with insights from 80,000+ expert users

Cloudera Distribution for Hadoop vs HPE Ezmeral Data Fabric comparison

 

Comparison Buyer's Guide

Executive Summary

Review summaries and opinions

We asked business professionals to review the solutions they use. Here are some excerpts of what they said:
 

Categories and Ranking

Cloudera Distribution for H...
Ranking in Hadoop
2nd
Average Rating
8.0
Reviews Sentiment
6.4
Number of Reviews
50
Ranking in other categories
NoSQL Databases (8th)
HPE Ezmeral Data Fabric
Ranking in Hadoop
4th
Average Rating
8.0
Reviews Sentiment
6.1
Number of Reviews
12
Ranking in other categories
No ranking in other categories
 

Mindshare comparison

As of March 2025, in the Hadoop category, the mindshare of Cloudera Distribution for Hadoop is 25.6%, up from 22.7% compared to the previous year. The mindshare of HPE Ezmeral Data Fabric is 14.7%, up from 10.4% compared to the previous year. It is calculated based on PeerSpot user engagement data.
Hadoop
 

Featured Reviews

Rok Dolinsek - PeerSpot reviewer
Enables on-premise implementation with powerful data processing capabilities
This is the only solution that is possible to install on-premise. Cloudera provides a hybrid solution that combines compute on cloud or on-premises. It includes all machine learning algorithms in the Spark machine learning library. All functionalities needed for a big data platform and ETL are on the platform, eliminating the need for other tools. It is scalable, ready for vertical scaling, and very powerful, offering numerous functionalities and configurations for generative AI.
Arnab Chatterjee - PeerSpot reviewer
It's flexible and easily accessible across multiple locations, but the upgrade process is complicated
Upgrading Ezmeral to a new version is a pain. They're trying to make the solution more container-friendly, so I think they're going in the right direction. The only problem we've had in the past was the upgrades. The process isn't smooth due to how the Red Hat operating system upgrades currently work. They're transforming their host stack to increase cloud readiness and edge compute capability. HPE is transitioning from a standard data-driven approach to one powered by AI analytics. That's something they have released very recently. I haven't tried that, but it will probably make things easier. The ability to adapt Ezmeral to the public cloud is probably missing. I've heard that they're getting leaner. However, it doesn't have a clear managed services offering for you if you want to deploy this stack on the cloud. That's a problem. This probably won't meet your needs if you require consistency across on-prem and the cloud. It's not Ezmeral's fault. None of the products would fit the bill. Cloud offerings are biased towards their own implementation. It's a general issue on most big data platforms. They're already working towards that, but it hasn't been released.

Quotes from Members

We asked business professionals to review the solutions they use. Here are some excerpts of what they said:
 

Pros

"I don't see any performance issues."
"This is the only solution that is possible to install on-premise."
"Provides a viable open-source solution for enterprise implementations and reliable, intelligent data analysis."
"We had a data warehouse before all the data. We can process a lot more data structures."
"The most valuable feature is that I can use CDH for almost all use cases across all industries, including the financial sector, public sector, private retailers, and so on."
"The tool's most interesting features are the distributed file system and unstructured data processing capability. Because we have a lot of unstructured data, like XML and social media logs, these features make it more valuable than the usual data warehousing solutions."
"The product provides better data processing features than other tools."
"The product is completely secure."
"HPE Ezmeral Data Fabric can be accessed from any namespace globally as you would access it from a machine using an NFS."
"It is a stable solution...It is a scalable solution."
"My customers find the product cheaper compared to other solutions. The previous solution that we used did not have unified analytics like the runtime or the analog."
"I like the administration part."
"The model creation was very interesting, especially with the libraries provided by the platform."
 

Cons

"It could be faster and more user-friendly."
"Cloudera's support is extremely bad and cannot be relied on."
"The competitors provide better functionalities."
"The solution is not fit for on-premise distributions."
"The areas of improvement depend on the scale of the project. For banking customers, security features and an essential budget for commercial licenses would be the top priority. Data regulation could be the most crucial for a project with extensive data or an extra use case."
"The solution does not support multiple languages very well and this means users need to create work-arounds to implement some solutions."
"The security of this solution could be improved. There should also be a way to basically have a blockchain enabled storage with the HDFS."
"The tool doesn't support reporting, and relational databases are still the major source of reporting data. Apache Iceberg will be launched soon within the Cloudera cluster for analytical purposes. The Cloudera Machine Learning aspect could be tuned and enhanced to enable us to host some predictive analytics machine learning and AI use cases."
"The product is not user-friendly."
"Having the ability to extend the services provided by the platform to an API architecture, a micro-services architecture, could be very helpful."
"The deployment could be faster. I want more support for the data lake in the next release."
"HPE Ezmeral Data Fabric is not compatible with third-party tools."
"Upgrading Ezmeral to a new version is a pain. They're trying to make the solution more container-friendly, so I think they're going in the right direction. The only problem we've had in the past was the upgrades. The process isn't smooth due to how the Red Hat operating system upgrades currently work."
 

Pricing and Cost Advice

"When comparing with Oracle Sybase and SQL, it's cheaper. It's not expensive."
"Cloudera Distribution for Hadoop is expensive, with support costs involved."
"The price is very high. The solution is expensive."
"The pricing must be improved."
"The solution is fairly expensive."
"It is an expensive product."
"The product’s price depends from project to project."
"Cloudera requires a license to use."
"The tool's price is cheap and based on a usage basis. The solution's licensing costs are yearly and there are no extra costs."
"HPE is flexible with you if you are an existing customer. They offer different models that might be beneficial for your organization. It all depends on how you negotiate."
"There is a need for my company to pay for the licensing costs of the solution."
report
Use our free recommendation engine to learn which Hadoop solutions are best for your needs.
839,319 professionals have used our research since 2012.
 

Top Industries

By visitors reading reviews
Financial Services Firm
23%
Computer Software Company
14%
Educational Organization
12%
Manufacturing Company
8%
Financial Services Firm
19%
Computer Software Company
14%
Retailer
8%
Government
7%
 

Company Size

By reviewers
Large Enterprise
Midsize Enterprise
Small Business
 

Questions from the Community

What do you like most about Cloudera Distribution for Hadoop?
The tool can be deployed using different container technologies, which makes it very scalable.
What is your experience regarding pricing and costs for Cloudera Distribution for Hadoop?
The price for Cloudera is average, yet it is very good compared to other solutions. It can be deployed on-premises, unlike competitors' cloud-only solutions.
What needs improvement with Cloudera Distribution for Hadoop?
It is quite complicated to configure and install. Integrating the platform into an information system is always a challenge, especially when starting with on-premise implementation. Integrating wit...
What do you like most about HPE Ezmeral Data Fabric?
It is a stable solution...It is a scalable solution.
What needs improvement with HPE Ezmeral Data Fabric?
There are some drawbacks in HPE Ezmeral Data Fabric when it comes to the interoperability part. HPE Ezmeral Data Fabric is not compatible with third-party tools. For example, HPE Ezmeral Data Fabri...
What is your primary use case for HPE Ezmeral Data Fabric?
The main purpose of HPE Ezmeral Data Fabric for me is that it acts as a database. In my company, we store our data with the help of HPE Ezmeral Data Fabric. It is possible to use Spark engine with ...
 

Also Known As

No data available
MapR, MapR Data Platform
 

Overview

 

Sample Customers

37signals, Adconion,adgooroo, Aggregate Knowledge, AMD, Apollo Group, Blackberry, Box, BT, CSC
Valence Health, Goodgame Studios, Pico, Terbium Labs, sovrn, Harte Hanks, Quantium, Razorsight, Novartis, Experian, Dentsu ix, Pontis Transitions, DataSong, Return Path, RAPP, HP
Find out what your peers are saying about Cloudera Distribution for Hadoop vs. HPE Ezmeral Data Fabric and other solutions. Updated: March 2025.
839,319 professionals have used our research since 2012.