Apache Cassandra vs. Apache Hadoop

Apache Cassandra

Apache Cassandra

95 Reviews and Ratings

Apache Hadoop

Apache Hadoop

270 Reviews and Ratings

Overview
Product	Rating	Most Used By	Product Summary	Starting Price
Cassandra	Score 9.0 out of 10	N/A	Cassandra is a no-SQL database from Apache.	N/A
Hadoop	Score 7.5 out of 10	N/A	Hadoop is an open source software from Apache, supporting distributed processing and data storage. Hadoop is popular for its scalability, reliability, and functionality available across commoditized hardware.	N/A

Pricing

Apache Cassandra

Apache Hadoop

Editions & Modules

No answers on this topic

No answers on this topic

Offerings

Pricing Offerings
Cassandra	Hadoop
Free Trial
No	No
Free/Freemium Version
No	Yes
Premium Consulting/Integration Services
No	No

Entry-level Setup Fee

No setup fee

No setup fee

Additional Details

—

—

More Pricing Information

Community Pulse
	Apache Cassandra	Apache Hadoop
Considered Both Products	Cassandra yixiang Shan IT Strategic Technical Advisor Chose Apache Cassandra We evaluated MongoDB also, but don't like the single point failure possibility. The HBase coupled us too tightly to the Hadoop world while we prefer more technical flexibility. Also HBase is designed for "cold"/old historical data lake use cases and is not typically used for … Incentivized Helpful? Rekha Joshi Staff Software Engineer Chose Apache Cassandra Apache Cassandra has the best of both worlds, it is a Java based NoSQL, linearly scalable, best in class tunable performance across different workloads, fault tolerant, distributed, masterless, time series database. We have used both Apache HBase and MongoDB for some use cases … Incentivized Helpful? David Prinzing Chief Technology Officer Chose Apache Cassandra Four years ago, I needed to choose a web-scale database. Having used relational databases for years (PostgreSQL is my favorite), I needed something that could perform well at scale with no downtime. I considered VoltDB for its in-memory speed, but it's limited in scale. I … Incentivized Helpful? Gary Ogasawara VP Engineering Chose Apache Cassandra Cassandra is well suited to more complex networks like multiple data centers. The underlying distributed systems logic is fundamentally sound. Incentivized Helpful?	Hadoop Gene Baker Vice President, Chief Architect, Development Manager and Software Engineer Chose Apache Hadoop Hands down, Hadoop is less expensive than the other platforms we considered. Cloudera was easier to set up but the expense ruled it out. MS-SQL didn't have the performance we saw with the Hadoop clusters and was more expensive. We considered MS-SQL mainly for its ability … Incentivized Helpful?

Features

Apache Cassandra

Apache Hadoop

NoSQL Databases

Comparison of NoSQL Databases features of Product A and Product B
	Apache Cassandra 8.0 5 Ratings 11% below category average	Apache Hadoop - Ratings
Performance	8.55 Ratings	00 Ratings
Availability	8.85 Ratings	00 Ratings
Concurrency	7.65 Ratings	00 Ratings
Security	8.05 Ratings	00 Ratings
Scalability	9.55 Ratings	00 Ratings
Data model flexibility	6.75 Ratings	00 Ratings
Deployment model flexibility	7.05 Ratings	00 Ratings

Best Alternatives
	Apache Cassandra	Apache Hadoop
Small Businesses	IBM Cloudant Score 7.4 out of 10	No answers on this topic
Medium-sized Companies	IBM Cloudant Score 7.4 out of 10	Cloudera Manager Score 9.9 out of 10
Enterprises	IBM Cloudant Score 7.4 out of 10	IBM Analytics Engine Score 7.1 out of 10
All Alternatives	View all alternatives	View all alternatives

User Ratings
	Apache Cassandra	Apache Hadoop
Likelihood to Recommend	6.0 (16 ratings)	8.0 (37 ratings)
Likelihood to Renew	8.6 (16 ratings)	9.6 (8 ratings)
Usability	7.0 (1 ratings)	8.0 (6 ratings)
Performance	- (0 ratings)	8.0 (1 ratings)
Support Rating	7.0 (1 ratings)	7.5 (3 ratings)
Online Training	- (0 ratings)	6.1 (2 ratings)
Implementation Rating	7.0 (1 ratings)	- (0 ratings)

User Testimonials
	Apache Cassandra	Apache Hadoop
Likelihood to Recommend	Apache Apache Cassandra is a NoSQL database and well suited where you need highly available, linearly scalable, tunable consistency and high performance across varying workloads. It has worked well for our use cases, and I shared my experiences to use it effectively at the last Cassandra summit! http://bit.ly/1Ok56TK It is a NoSQL database, finally you can tune it to be strongly consistent and successfully use it as such. However those are not usual patterns, as you negotiate on latency. It works well if you require that. If your use case needs strongly consistent environments with semantics of a relational database or if the use case needs a data warehouse, or if you need NoSQL with ACID transactions, Apache Cassandra may not be the optimum choice. Incentivized Rekha Joshi Staff Software Engineer Read full review	Apache Altogether, I want to say that Apache Hadoop is well-suited to a larger and unstructured data flow like an aggregation of web traffic or even advertising. I think Apache Hadoop is great when you literally have petabytes of data that need to be stored and processed on an ongoing basis. Also, I would recommend that the software should be supplemented with a faster and interactive database for a better querying service. Lastly, it's very cost-effective so it is good to give it a shot before coming to any conclusion. Incentivized Peter Suter Senior Software Engineer (GUI) Read full review
Pros	Apache Continuous availability: as a fully distributed database (no master nodes), we can update nodes with rolling restarts and accommodate minor outages without impacting our customer services. Linear scalability: for every unit of compute that you add, you get an equivalent unit of capacity. The same application can scale from a single developer's laptop to a web-scale service with billions of rows in a table. Amazing performance: if you design your data model correctly, bearing in mind the queries you need to answer, you can get answers in milliseconds. Time-series data: Cassandra excels at recording, processing, and retrieving time-series data. It's a simple matter to version everything and simply record what happens, rather than going back and editing things. Then, you can compute things from the recorded history. Incentivized David Prinzing Chief Technology Officer Read full review	Apache Handles large amounts of unstructured data well, for business level purposes Is a good catchall because of this design, i.e. what does not fit into our vertical tables fits here. Decent for large ETL pipelines and logging free-for-alls because of this, also. Incentivized JH Joe Hughes Senior DevOps Engineer Read full review
Cons	Apache Cassandra runs on the JVM and therefor may require a lot of GC tuning for read/write intensive applications. Requires manual periodic maintenance - for example it is recommended to run a cleanup on a regular basis. There are a lot of knobs and buttons to configure the system. For many cases the default configuration will be sufficient, but if its not - you will need significant ramp up on the inner workings of Cassandra in order to effectively tune it. Incentivized Verified User Anonymous Read full review	Apache Less organizational support system. Bugs need to be fixed and outside help take a long time to push updates Not for small data sets Data security needs to be ramped up Failure in NameNode has no replication which takes a lot of time to recover Incentivized Bharadwaj (Brad) Chivukula Sr. Engineering Manager/Delivery Manager Read full review
Likelihood to Renew	Apache I would recommend Cassandra DB to those who know their use case very well, as well as know how they are going to store and retrieve data. If you need a guarantee in data storage and retrieval, and a DB that can be linearly grown by adding nodes across availability zones and regions, then this is the database you should choose. Incentivized Verified User Anonymous Read full review	Apache Hadoop is organization-independent and can be used for various purposes ranging from archiving to reporting and can make use of economic, commodity hardware. There is also a lot of saving in terms of licensing costs - since most of the Hadoop ecosystem is available as open-source and is free Bhushan Lakhe Senior Vice President Read full review
Usability	Apache It’s great tool but it can be complicated when it comes administration and maintenance. Incentivized Glen Kim Senior Software Engineer Read full review	Apache As Hadoop enterprise licensed version is quite fine tuned and easy to use makes it good choice for Hadoop administrators. It’s scalability and integration with Kerberos is good option for authentication and authorisation. installation can be improved. logging can be improved so that it become easier for debugging purposes. parallel processing of data is achieved easily. Incentivized Verified User Anonymous Read full review
Support Rating	Apache Sometimes instead giving straight answer, we ‘re getting transfered to talk professional service. Incentivized Glen Kim Senior Software Engineer Read full review	Apache It's a great value for what you pay, and most Data Base Administrators (DBAs) can walk in and use it without substantial training. I tend to dabble on the analyst side, so querying the data I need feels like it can take forever, especially on higher traffic days like Monday. Incentivized Blake Baron Senior Financial Analyst Read full review
Online Training	Apache No answers on this topic	Apache Hadoop is a complex topic and best suited for classrom training. Online training are a waste of time and money. Bhushan Lakhe Senior Vice President Read full review
Alternatives Considered	Apache We evaluated MongoDB also, but don't like the single point failure possibility. The HBase coupled us too tightly to the Hadoop world while we prefer more technical flexibility. Also HBase is designed for "cold"/old historical data lake use cases and is not typically used for web and mobile applications due to its performance concern. Cassandra, by contrast, offers the availability and performance necessary for developing highly available applications. Furthermore, the Hadoop technology stack is typically deployed in a single location, while in the big international enterprise context, we demand the feasibility for deployment across countries and continents, hence finally we are favor of Cassandra Incentivized yixiang Shan IT Strategic Technical Advisor Read full review	Apache Not used any other product than Hadoop and I don't think our company will switch to any other product, as Hadoop is providing excellent results. Our company is growing rapidly, Hadoop helps to keep up our performance and meet customer expectations. We also use HDFS which provides very high bandwidth to support MapReduce workloads. Incentivized Verified User Anonymous Read full review
Return on Investment	Apache I have no experience with this but from the blogs and news what I believe is that in businesses where there is high demand for scalability, Cassandra is a good choice to go for. Since it works on CQL, it is quite familiar with SQL in understanding therefore it does not prevent a new employee to start in learning and having the Cassandra experience at an industrial level. Incentivized Verified User Anonymous Read full review	Apache There are many advantages of Hadoop as first it has made the management and processing of extremely colossal data very easy and has simplified the lives of so many people including me. Hadoop is quite interesting due to its new and improved features plus innovative functions. Incentivized Chantel Moreno Finance & Accounting Professional Read full review
ScreenShots