Azure HDInsight vs. IBM watsonx.data

Overview
ProductRatingMost Used ByProduct SummaryStarting Price
Azure HDInsight
Score 8.1 out of 10
N/A
HDInsight is an implementation of the Apache Hadoop technology stack on the Microsoft Azure cloud platform: It is based on the Hortonworks Hadoop distribution. Microsoft Azure HDInsight includes implementations of Apache Spark, HBase, Storm, Pig, Hive, Sqoop, Oozie, Ambari, etc. It also integrates with with business intelligence (BI) tools such as Power BI, Excel, SQL Server Analysis Services, and SQL Server Reporting Services.N/A
IBM watsonx.data
Score 8.8 out of 10
N/A
Watsonx.data is presented as an open, hybrid and governed data store that makes it possible for enterprises to scale analytics and AI with a fit-for-purpose data store, built on an open lakehouse architecture, supported by querying, governance and open data formats to access and share data.N/A
Pricing
Azure HDInsightIBM watsonx.data
Editions & Modules
No answers on this topic
No answers on this topic
Offerings
Pricing Offerings
Azure HDInsightIBM watsonx.data
Free Trial
NoYes
Free/Freemium Version
NoNo
Premium Consulting/Integration Services
NoNo
Entry-level Setup FeeNo setup feeNo setup fee
Additional Details
More Pricing Information
Community Pulse
Azure HDInsightIBM watsonx.data
Best Alternatives
Azure HDInsightIBM watsonx.data
Small Businesses

No answers on this topic

No answers on this topic

Medium-sized Companies
Cloudera Manager
Cloudera Manager
Score 9.9 out of 10
Snowflake
Snowflake
Score 8.7 out of 10
Enterprises
IBM Analytics Engine
IBM Analytics Engine
Score 7.2 out of 10
Snowflake
Snowflake
Score 8.7 out of 10
All AlternativesView all alternativesView all alternatives
User Ratings
Azure HDInsightIBM watsonx.data
Likelihood to Recommend
4.0
(6 ratings)
8.8
(27 ratings)
Likelihood to Renew
-
(0 ratings)
7.7
(3 ratings)
Usability
8.9
(4 ratings)
7.6
(9 ratings)
Support Rating
1.0
(5 ratings)
9.3
(3 ratings)
User Testimonials
Azure HDInsightIBM watsonx.data
Likelihood to Recommend
Microsoft
Well suited: A tiny-mid sized company with no immediate plans of growing the volume of their data processing, that can afford long response times from support. Also it helps if you are not prone to put your hands on Linux and Spark configuration. In fact, it can make things go really faster if you also work with the bundle-in Jupyter. And, if you need to perform some diagnostics and / or administrative tasks, that's full of tools to find an understand the Root Cause. Ideal for non experts. Less appropriate: Big Data company, intense on demand cluster creation, mission critical, costs reduction, latest versions of libraries required, sophisticate customizations required.
Read full review
IBM
Real-time transaction processing (both reads and writes) is where DataStax Enterprise shines. It's very fast with linear scalability should more resources be needed. Additional nodes are added very easily. DataStax Enterprise on its own (without Solr or Spark enabled) isn't well suited for long complicated reports. The data model doesn't support joining multiple tables together which is common in BI reporting.
Read full review
Pros
Microsoft
  • Data is presented without interfering others (IT or other dept).
  • Data is managed properly and is available for retrievable any time.
  • Legacy use of CD/DVD and Pendrive are not required.
Read full review
IBM
  • Datastax Cassandra provides high availability and good performance for a database. It is built on top of open source Apache Cassandra so you can always somewhat understand the internal functioning and why.
  • Datastax Cassandra is fairly simple to start using, you can install/setup your cluster and be productive in 1 day.
  • Datastax Cassandra provides a lot of good detailed documentation, and when starting, the detailed free videos on the Datastax site and documentation are very helpful.
  • Datastax Enterprise Edition of Cassandra provides more tools, good support, and quick response SLA for enterprise business support.
Read full review
Cons
Microsoft
  • The only problem I have come across is when loading large volumes of data I sometimes get an error message, I assume this means something is corrupt from within. I would love a way for this to be resolved without having to start over.
Read full review
IBM
  • Integration complexity with Security Tools while watsonx.Data is well-suited for native tools, but integration with third-party security tools requires custom connectors or manual ETL pipelines. which leads to an increase in setup time.
  • User interface and query time can be improved.
Read full review
Likelihood to Renew
Microsoft
No answers on this topic
IBM
As an open source technology Cassandra can be readily used with or without any commercial support. DataStax provides value-added services and features, and in the end it is up to individual situations to strike a balance between the desirability of such support/service versus the associated cost.
Read full review
Usability
Microsoft
Azure HDInsight is usable on the top of Azure Data Lake and gives us the benefit of analyzing large scale data workload in Hadoop. Usability and support from Microsoft are outstanding.
Read full review
IBM
DataStax has a good community built around it and has amazing scalability options. Though the initial setup is a bit costly, in the long run, it makes up for it. It also has powerful monitoring tools and a clean UI.
Read full review
Reliability and Availability
Microsoft
No answers on this topic
IBM
good recovery features
Read full review
Performance
Microsoft
No answers on this topic
IBM
scalable product
Read full review
Support Rating
Microsoft
Inexpert, isolated teams... not good for support an excessively complex platform. Lots of weeks or months for a complex problem troubleshoot. Many time lost stuck on MindTree, before the case was finally escalated with Microsoft!
Read full review
IBM
We have had a few situations where we caused an outage or something has gone wrong and we are able to get a support person to offer live help within minutes. The escalation process is excellent - the best I've seen - and the support team is incredibly strong. Outside of emergencies, the team is very helpful with general questions and working through data model exercises and the subscription I believe still comes with some hours to help get the data model reviewed.
Read full review
Online Training
Microsoft
No answers on this topic
IBM
easy to follow documentation, support is there when needed
Read full review
Implementation Rating
Microsoft
No answers on this topic
IBM
use saas service
Read full review
Alternatives Considered
Microsoft
At this time I have not used any other similar products... I am open to it but Azure HDInsight and its components really work well for our organization.
Read full review
IBM
Pinecone and IBM watsonx.data (Milvus in our case) both work great as a full-managed cloud-based vector database. We selected IBM watsonx.data because it integrates well with watson.ai and is a little more beginner friendly than Pinecone, but I think both are great anyway.
Read full review
Scalability
Microsoft
No answers on this topic
IBM
cognos integration works great
Read full review
Return on Investment
Microsoft
  • ROI is of course there, as no legacy software for data presentation.
  • No manual intervention for data retrieval.
  • Data is available anywhere as requested.
Read full review
IBM
  • for one automation project, we managed to cut cloud storage costs by a third through IBM watsonx.data's lakehouse optimization
  • data integration projects have had a 20 % reduction in turnaround times. Can only imagine how that will improve with the Claude partnership
Read full review
ScreenShots