Apache Cassandra vs. Apache Hadoop vs. PostgreSQL

Apache Cassandra

Apache Cassandra

95 Reviews and Ratings

Apache Hadoop

Apache Hadoop

270 Reviews and Ratings

PostgreSQL

PostgreSQL

354 Reviews and Ratings

Overview
Product	Rating	Most Used By	Product Summary	Starting Price
Cassandra	Score 8.9 out of 10	N/A	Cassandra is a no-SQL database from Apache.	N/A
Hadoop	Score 7.5 out of 10	N/A	Hadoop is an open source software from Apache, supporting distributed processing and data storage. Hadoop is popular for its scalability, reliability, and functionality available across commoditized hardware.	N/A
PostgreSQL	Score 8.7 out of 10	N/A	PostgreSQL (alternately Postgres) is a free and open source object-relational database system boasting over 30 years of active development, reliability, feature robustness, and performance. It supports SQL and is designed to support various workloads flexibly.	N/A

Pricing

Apache Cassandra

Apache Hadoop

PostgreSQL

Editions & Modules

No answers on this topic

No answers on this topic

No answers on this topic

Offerings

Pricing Offerings
Cassandra	Hadoop	PostgreSQL
Free Trial
No	No	No
Free/Freemium Version
No	Yes	No
Premium Consulting/Integration Services
No	No	No

Entry-level Setup Fee

No setup fee

No setup fee

No setup fee

Additional Details

—

—

—

More Pricing Information

Community Pulse
	Apache Cassandra	Apache Hadoop	PostgreSQL
Considered Multiple Products	Cassandra David Prinzing Chief Technology Officer Chose Apache Cassandra Four years ago, I needed to choose a web-scale database. Having used relational databases for years (PostgreSQL is my favorite), I needed something that could perform well at scale with no downtime. I considered VoltDB for its in-memory speed, but it's limited in scale. I … Incentivized Helpful? Verified User Team Lead Chose Apache Cassandra DynamoDB is good and is also a truly global database as a service on AWS. However, if your organization is not using AWS, then Cassandra will provide a highly scalable and tuneable, consistent database. Cassandra is also fault-tolerant and good for replication across multiple … Incentivized Helpful? yixiang Shan IT Strategic Technical Advisor Chose Apache Cassandra We evaluated MongoDB also, but don't like the single point failure possibility. The HBase coupled us too tightly to the Hadoop world while we prefer more technical flexibility. Also HBase is designed for "cold"/old historical data lake use cases and is not typically used for … Incentivized Helpful? Verified User Engineer Chose Apache Cassandra Cassandra is the only NoSQL database I have extensive experience with. In terms of other open source database solutions, I can say that I like Cassandra as much or equally as traditional Oracle MySQL, and a lot more than PostgresSQL. The decision to use Cassandra was driven by … Incentivized Helpful? Rekha Joshi Staff Software Engineer Chose Apache Cassandra Apache Cassandra has the best of both worlds, it is a Java based NoSQL, linearly scalable, best in class tunable performance across different workloads, fault tolerant, distributed, masterless, time series database. We have used both Apache HBase and MongoDB for some use cases … Incentivized Helpful? Gary Ogasawara VP Engineering Chose Apache Cassandra Cassandra is well suited to more complex networks like multiple data centers. The underlying distributed systems logic is fundamentally sound. Incentivized Helpful?	Hadoop Gene Baker Vice President, Chief Architect, Development Manager and Software Engineer Chose Apache Hadoop Hands down, Hadoop is less expensive than the other platforms we considered. Cloudera was easier to set up but the expense ruled it out. MS-SQL didn't have the performance we saw with the Hadoop clusters and was more expensive. We considered MS-SQL mainly for its ability … Incentivized Helpful? Verified User Engineer Chose Apache Hadoop I haven't worked with other Big Data aggregation services like Hadoop. As far as I know, Hadoop is the leading choice in this field with good cause. There is a lot of community support, custom modules, paid consultants, free and paid training. All this makes it an ideal choice … Incentivized Helpful?	PostgreSQL Ron Ballard Consultant and author Chose PostgreSQL In my experience using all of these products over many years, PostgreSQL is better than any of them in reliability, performance, productivity, cost, scalability and interoperability across operating systems. Incentivized Helpful? Anson Abraham Data Czar Chose PostgreSQL PostgrPostgreSQL as a transaction db engine against oracle and sql server works well. TPM wise compared to MySQL and MariaDB, on an evan scale. SQL function supports, far outweighs compared to MySQL and MariaDB. PG Extensions allow for flexibiltity and scalability. Allows … Incentivized Helpful? Verified User Administrator Chose PostgreSQL As I have been telling all along, PostgreSQL is much cheaper compared to the other RDBMS solutions. It has got better performance with some of the application services that we are using and is easy to maintain. Overall, we are satisfied migrating to PostgreSQL database clusters. Incentivized Helpful? Balázs Kiss Software Developer Chose PostgreSQL It's a viable alternative, with a rich feature set and a reliable system. PostgreSQL is one of the best RDBMS's currently on the market in 2020, it serves just as well as a starter, PoC DB for any software idea as a final, highly valuable database solution for big systems. Incentivized Helpful?

Features

Apache Cassandra

Apache Hadoop

PostgreSQL

NoSQL Databases

Comparison of NoSQL Databases features of Product A and Product B
	Apache Cassandra 8.0 5 Ratings 11% below category average	Apache Hadoop - Ratings	PostgreSQL - Ratings
Performance	8.55 Ratings	00 Ratings	00 Ratings
Availability	8.85 Ratings	00 Ratings	00 Ratings
Concurrency	7.65 Ratings	00 Ratings	00 Ratings
Security	8.05 Ratings	00 Ratings	00 Ratings
Scalability	9.55 Ratings	00 Ratings	00 Ratings
Data model flexibility	6.75 Ratings	00 Ratings	00 Ratings
Deployment model flexibility	7.05 Ratings	00 Ratings	00 Ratings

Best Alternatives
	Apache Cassandra	Apache Hadoop	PostgreSQL
Small Businesses	IBM Cloudant Score 7.4 out of 10	No answers on this topic	InfluxDB Score 8.8 out of 10
Medium-sized Companies	IBM Cloudant Score 7.4 out of 10	Cloudera Manager Score 9.9 out of 10	SQLite Score 8.0 out of 10
Enterprises	IBM Cloudant Score 7.4 out of 10	IBM Analytics Engine Score 7.2 out of 10	SQLite Score 8.0 out of 10
All Alternatives	View all alternatives	View all alternatives	View all alternatives

User Ratings
	Apache Cassandra	Apache Hadoop	PostgreSQL
Likelihood to Recommend	6.0 (16 ratings)	8.0 (37 ratings)	8.0 (55 ratings)
Likelihood to Renew	8.6 (16 ratings)	9.6 (8 ratings)	9.0 (1 ratings)
Usability	7.0 (1 ratings)	8.0 (6 ratings)	8.3 (9 ratings)
Availability	- (0 ratings)	- (0 ratings)	9.0 (1 ratings)
Performance	- (0 ratings)	8.0 (1 ratings)	7.0 (1 ratings)
Support Rating	7.0 (1 ratings)	7.5 (3 ratings)	9.3 (7 ratings)
Online Training	- (0 ratings)	6.1 (2 ratings)	- (0 ratings)
Implementation Rating	7.0 (1 ratings)	- (0 ratings)	9.0 (1 ratings)
Product Scalability	- (0 ratings)	- (0 ratings)	8.0 (1 ratings)

User Testimonials
	Apache Cassandra	Apache Hadoop	PostgreSQL
Likelihood to Recommend	Apache Apache Cassandra is a NoSQL database and well suited where you need highly available, linearly scalable, tunable consistency and high performance across varying workloads. It has worked well for our use cases, and I shared my experiences to use it effectively at the last Cassandra summit! http://bit.ly/1Ok56TK It is a NoSQL database, finally you can tune it to be strongly consistent and successfully use it as such. However those are not usual patterns, as you negotiate on latency. It works well if you require that. If your use case needs strongly consistent environments with semantics of a relational database or if the use case needs a data warehouse, or if you need NoSQL with ACID transactions, Apache Cassandra may not be the optimum choice. Incentivized Rekha Joshi Staff Software Engineer Read full review	Apache Altogether, I want to say that Apache Hadoop is well-suited to a larger and unstructured data flow like an aggregation of web traffic or even advertising. I think Apache Hadoop is great when you literally have petabytes of data that need to be stored and processed on an ongoing basis. Also, I would recommend that the software should be supplemented with a faster and interactive database for a better querying service. Lastly, it's very cost-effective so it is good to give it a shot before coming to any conclusion. Incentivized Peter Suter Senior Software Engineer (GUI) Read full review	PostgreSQL Global Development Group PostgreSQL is best used for structured data, and best when following relational database design principles. I would not use PostgreSQL for large unstructured data such as video, images, sound files, xml documents, web-pages, especially if these files have their own highly variable, internal structure. Incentivized Ron Ballard Consultant and author Read full review
Pros	Apache Continuous availability: as a fully distributed database (no master nodes), we can update nodes with rolling restarts and accommodate minor outages without impacting our customer services. Linear scalability: for every unit of compute that you add, you get an equivalent unit of capacity. The same application can scale from a single developer's laptop to a web-scale service with billions of rows in a table. Amazing performance: if you design your data model correctly, bearing in mind the queries you need to answer, you can get answers in milliseconds. Time-series data: Cassandra excels at recording, processing, and retrieving time-series data. It's a simple matter to version everything and simply record what happens, rather than going back and editing things. Then, you can compute things from the recorded history. Incentivized David Prinzing Chief Technology Officer Read full review	Apache Handles large amounts of unstructured data well, for business level purposes Is a good catchall because of this design, i.e. what does not fit into our vertical tables fits here. Decent for large ETL pipelines and logging free-for-alls because of this, also. Incentivized JH Joe Hughes Senior DevOps Engineer Read full review	PostgreSQL Global Development Group It works well with external data sources and runs on platforms with stable performance. Clients can rest assured that their personal information will be safe and secure. Many forums discuss setup and usage, and most are free. Adding tooling applications to a computer is unlimited. PostgreSQL runs on many OS platforms and supports ANSI SQL, stored procedures, and triggers. Aurpa Fiza Software Application Developer Read full review
Cons	Apache Cassandra runs on the JVM and therefor may require a lot of GC tuning for read/write intensive applications. Requires manual periodic maintenance - for example it is recommended to run a cleanup on a regular basis. There are a lot of knobs and buttons to configure the system. For many cases the default configuration will be sufficient, but if its not - you will need significant ramp up on the inner workings of Cassandra in order to effectively tune it. Incentivized Verified User Anonymous Read full review	Apache Less organizational support system. Bugs need to be fixed and outside help take a long time to push updates Not for small data sets Data security needs to be ramped up Failure in NameNode has no replication which takes a lot of time to recover Incentivized Bharadwaj (Brad) Chivukula Sr. Engineering Manager/Delivery Manager Read full review	PostgreSQL Global Development Group Clearer indications on what is the query plan, to optimize the query More out of the box, Postgres specific, SQL functions It would be nice to have a more visual aid of the relationship between all tables, but possibly this depend more on the UI used Incentivized Verified User Anonymous Read full review
Likelihood to Renew	Apache I would recommend Cassandra DB to those who know their use case very well, as well as know how they are going to store and retrieve data. If you need a guarantee in data storage and retrieval, and a DB that can be linearly grown by adding nodes across availability zones and regions, then this is the database you should choose. Incentivized Verified User Anonymous Read full review	Apache Hadoop is organization-independent and can be used for various purposes ranging from archiving to reporting and can make use of economic, commodity hardware. There is also a lot of saving in terms of licensing costs - since most of the Hadoop ecosystem is available as open-source and is free Bhushan Lakhe Senior Vice President Read full review	PostgreSQL Global Development Group As a needed software for day to day development activities Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review
Usability	Apache It’s great tool but it can be complicated when it comes administration and maintenance. Incentivized Glen Kim Senior Software Engineer Read full review	Apache As Hadoop enterprise licensed version is quite fine tuned and easy to use makes it good choice for Hadoop administrators. It’s scalability and integration with Kerberos is good option for authentication and authorisation. installation can be improved. logging can be improved so that it become easier for debugging purposes. parallel processing of data is achieved easily. Incentivized Verified User Anonymous Read full review	PostgreSQL Global Development Group Postgresql is the best tool out there for relational data so I have to give it a high rating when it comes to analytics, data availability and consistency, so on and so forth. SQL is also a relatively consistent language so when it comes to building new tables and loading data in from the OLTP database, there are enough tools where we can perform ETL on a scalable basis. Incentivized Verified User Anonymous Read full review
Reliability and Availability	Apache No answers on this topic	Apache No answers on this topic	PostgreSQL Global Development Group PostgreSQL's availability is top notch. Apart from connection time-out for an idle user, the database is super reliable. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review
Performance	Apache No answers on this topic	Apache No answers on this topic	PostgreSQL Global Development Group The data queries are relatively quick for a small to medium sized table. With complex joins, and a wide and deep table however, the performance of the query has room for improvement. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review
Support Rating	Apache Sometimes instead giving straight answer, we ‘re getting transfered to talk professional service. Incentivized Glen Kim Senior Software Engineer Read full review	Apache It's a great value for what you pay, and most Data Base Administrators (DBAs) can walk in and use it without substantial training. I tend to dabble on the analyst side, so querying the data I need feels like it can take forever, especially on higher traffic days like Monday. Incentivized Blake Baron Senior Financial Analyst Read full review	PostgreSQL Global Development Group There are several companies that you can contract for technical support, like EnterpriseDB or Percona, both first level in expertise and commitment to the software. But we do not have contracts with them, we have done all the way from googling to forums, and never have a problem that we cannot resolve or pass around. And for dozens of projects and more than 15 years now. Incentivized Javier Blanque Technology Risk and Information Assets Manager Read full review
Online Training	Apache No answers on this topic	Apache Hadoop is a complex topic and best suited for classrom training. Online training are a waste of time and money. Bhushan Lakhe Senior Vice President Read full review	PostgreSQL Global Development Group The online training is request based. Had there been recorded videos available online for potential users to benefit from, I could have rated it higher. The online documentation however is very helpful. The online documentation PDF is downloadable and allows users to pace their own learning. With examples and code snippets, the documentation is great starting point. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review
Implementation Rating	Apache No answers on this topic	Apache No answers on this topic	PostgreSQL Global Development Group The online documentation of the PostgreSQL product is elaborate and takes users step by step. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review
Alternatives Considered	Apache We evaluated MongoDB also, but don't like the single point failure possibility. The HBase coupled us too tightly to the Hadoop world while we prefer more technical flexibility. Also HBase is designed for "cold"/old historical data lake use cases and is not typically used for web and mobile applications due to its performance concern. Cassandra, by contrast, offers the availability and performance necessary for developing highly available applications. Furthermore, the Hadoop technology stack is typically deployed in a single location, while in the big international enterprise context, we demand the feasibility for deployment across countries and continents, hence finally we are favor of Cassandra Incentivized yixiang Shan IT Strategic Technical Advisor Read full review	Apache Not used any other product than Hadoop and I don't think our company will switch to any other product, as Hadoop is providing excellent results. Our company is growing rapidly, Hadoop helps to keep up our performance and meet customer expectations. We also use HDFS which provides very high bandwidth to support MapReduce workloads. Incentivized Verified User Anonymous Read full review	PostgreSQL Global Development Group Although the competition between the different databases is increasingly aggressive in the sense that they provide many improvements, new functionalities, compatibility with complementary components or environments, in some cases it requires that it be followed within the same family of applications that performs the company that develops it and that is not all bad, but being able to adapt or configure different programs, applications or other environments developed by third parties apart is what gives PostgreSQL a certain advantage and this diversification in the components that can be joined with it, is the reason why it is a great option to choose. Incentivized Moris Mendez Ing. de Sistemas Informaticos Read full review
Scalability	Apache No answers on this topic	Apache No answers on this topic	PostgreSQL Global Development Group The DB is reliable, scalable, easy to use and resolves most DB needs Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review
Return on Investment	Apache I have no experience with this but from the blogs and news what I believe is that in businesses where there is high demand for scalability, Cassandra is a good choice to go for. Since it works on CQL, it is quite familiar with SQL in understanding therefore it does not prevent a new employee to start in learning and having the Cassandra experience at an industrial level. Incentivized Verified User Anonymous Read full review	Apache There are many advantages of Hadoop as first it has made the management and processing of extremely colossal data very easy and has simplified the lives of so many people including me. Hadoop is quite interesting due to its new and improved features plus innovative functions. Incentivized Chantel Moreno Finance & Accounting Professional Read full review	PostgreSQL Global Development Group Easy to administer so our DevOps team has only ever used minimal time to setup, tune, and maintain. Easy to interface with so our Engineering team has only ever used minimal time to query or modify the database. Getting the data is straightforward, what we do with it is the bigger concern. It's free. You can't beat that. Incentivized Don Burks Technical Lead Read full review
ScreenShots