We compared IBM InfoSphere DataStage and IBM Cloud Pak for Data based on our user's reviews in several parameters.
IBM InfoSphere DataStage is praised for its strong data integration, connectors, workflow management, ETL functionalities, and data quality controls. In contrast, IBM Cloud Pak for Data is commended for its analytics capabilities, user interface, data management tools, integration, scalability, governance, security, collaboration, and AI-driven features. Feedback on customer service, setup duration, pricing, and ROI varies between the two products.
Features: IBM InfoSphere DataStage is praised for its strong data integration capabilities, comprehensive set of connectors, efficient workflow management, and robust ETL functionalities. On the other hand, IBM Cloud Pak for Data is valued for its robust analytics capabilities, ease of use, comprehensive data management tools, seamless integration, and advanced data governance and security features. It also offers AI-driven capabilities like machine learning and predictive analytics.
Pricing and ROI: The available data does not provide any information about the setup cost for IBM InfoSphere DataStage. Similarly, the pricing and licensing information for IBM Cloud Pak for Data is not provided in the available data source., IBM InfoSphere DataStage has no available data to determine its ROI, while there is also no information or insights about the ROI of IBM Cloud Pak for Data.
Room for Improvement: IBM InfoSphere DataStage does not have specific areas for improvement identified in the available responses. Similarly, there is no specific feedback or review available for IBM Cloud Pak for Data on what needs improvement.
Deployment and customer support: Based on the available summaries, it is not possible to compare the user reviews regarding the duration to establish IBM InfoSphere DataStage and IBM Cloud Pak for Data as the feedback related to these aspects is not provided for both products., Based on the available data, there is not enough information to provide a summary of the customer service and support of IBM InfoSphere DataStage. The customer service and support of IBM Cloud Pak for Data received a lack of feedback from the reviews provided.
The summary above is based on 24 interviews we conducted recently with IBM InfoSphere DataStage and IBM Cloud Pak for Data users. To access the review's full transcripts, download our report.
"One of Cloud Pak's best features is the Watson Knowledge Catalog, which helps you implement data governance."
"Scalability-wise, I rate the solution a nine or ten out of ten."
"What I found most helpful in IBM Cloud Pak for Data is containerization, which means it's easy to shift and leave in terms of moving to other clouds. That's an advantage of IBM Cloud Pak for Data."
"Its data preparation capabilities are highly valuable."
"Cloud Pak's most valuable features are IBM MQ, IBM App Connect, IBM API Connect, and ISPF."
"The most valuable features are data virtualization and reporting."
"You can model the data there, connect the data models with the business processes and create data lineage processes."
"The most valuable features of IBM Cloud Pak for Data are the Watson Studio, where we can initiate more groups and write code. Additionally, Watson Machine Learning is available with many other services, such as APIs which you can plug the machine learning models."
"The product is easy to deploy."
"ETL is the most valuable feature."
"The solution's scalability is really good...we are using multi-instance jobs where you can scale them easily."
"Finding logs is very easy on the solution."
"The most valuable feature is the data integration for data warehousing."
"The most valuable feature is the product's versatility to inject data."
"I am impressed with the tool's ETL tracing."
"When we have needed help from the IBM team, they were helpful. Our company is a premium partner so we get fast responses."
"Cloud Pak would be improved with integration with cloud service providers like Cloudera."
"The technical support could be a little better."
"The solution's user experience is an area that has room for improvement."
"One challenge I'm facing with IBM Cloud Pak for Data is native features have been decommissioned, such as XML input and output. Too many changes have been made, and my company has around one hundred thousand mappings, so my team has been putting more effort into alternative ways to do things. Another area for improvement in IBM Cloud Pak for Data is that it's more complicated to shift from on-premise to the cloud. Other vendors provide secure agents that easily connect with your existing setup. Still, with IBM Cloud Pak for Data, you have to perform connection migration steps, upgrade to the latest version, etc., which makes it more complicated, especially as my company has XML-based mappings. Still, the XML input and output capabilities of IBM Cloud Pak for Data have been discontinued, so I'd like IBM to bring that back."
"The solution could have more connectors."
"The interface could improve because sometimes it becomes slow. Sometimes there is a delay between clicks when using the software, which can make the development process slow. It can take a few seconds to complete one action, and then a few more seconds to do the next one."
"The tool depends on the control plane, an OpenShift container platform utilized as an orchestration layer...So, we have communicated this issue to IBM and asked if it is feasible to adapt the solution to work on a Kubernetes platform that we support."
"The product must improve its performance."
"The graphical user interface (GUI) feels a lot like the interfaces from the 1980s."
"The setup is extremely difficult."
"There are three things that could improve - the cloud, monitoring and cloud integration. It's a solid product but not a modern one and of course it depends what you're looking for."
"There could be more customization options for the product."
"It takes a lot of time to actually trigger your job and then go into the logs and other stuff. So all of this is really time-consuming."
"In terms of intermediate storage, we have some challenges, especially with customers who store data in intermediate locations."
"It would be useful to provide support for Python, AR, and Java."
"Their web interface is good but the on-prem sites are outdated. The solution could also be improved if they could integrate the data pipeline scheduling part of their interface."
IBM Cloud Pak for Data is ranked 16th in Data Integration with 11 reviews while IBM InfoSphere DataStage is ranked 7th in Data Integration with 37 reviews. IBM Cloud Pak for Data is rated 8.0, while IBM InfoSphere DataStage is rated 7.8. The top reviewer of IBM Cloud Pak for Data writes "A scalable data analytics and digital transformation tool that provides useful features and integrations". On the other hand, the top reviewer of IBM InfoSphere DataStage writes "User-friendly with a lot of functions for transmission rules, but has slow performance and not suitable for a huge volume of data". IBM Cloud Pak for Data is most compared with Azure Data Factory, Informatica Cloud Data Integration, Palantir Foundry, Denodo and IBM InfoSphere Information Server, whereas IBM InfoSphere DataStage is most compared with SSIS, Azure Data Factory, Talend Open Studio, Informatica PowerCenter and IBM InfoSphere Information Server. See our IBM Cloud Pak for Data vs. IBM InfoSphere DataStage report.
See our list of best Data Integration vendors.
We monitor all Data Integration reviews to prevent fraudulent reviews and keep review quality high. We do not post reviews by company employees or direct competitors. We validate each review for authenticity via cross-reference with LinkedIn, and personal follow-up with the reviewer when necessary.