Apache Spark vs. PostgreSQL vs. SAP HANA Cloud

Apache Spark

165 Reviews and Ratings

PostgreSQL

354 Reviews and Ratings

SAP HANA Cloud

1019 Reviews and Ratings

Learn More

Overview
Product	Rating	Most Used By	Product Summary	Starting Price
Apache Spark	Score 8.9 out of 10	N/A	Apache Spark is a multi-language engine for executing data engineering, data science, and machine learning on single-node machines or clusters.	N/A
PostgreSQL	Score 8.7 out of 10	N/A	PostgreSQL (alternately Postgres) is a free and open source object-relational database system boasting over 30 years of active development, reliability, feature robustness, and performance. It supports SQL and is designed to support various workloads flexibly.	N/A
SAP HANA Cloud	Score 8.9 out of 10	N/A	SAP HANA is an application that uses in-memory database technology to process very large amounts of real-time data from relational databases, both SAP and non-SAP, in a very short time. The in-memory computing engine allows HANA to process data stored in RAM as opposed to reading it from a disk which means that the data can be accessed in real time by the applications using HANA. The product is sold both as an appliance and as a cloud-based software solution.	$0.95 per month Capacity Units

Pricing

Apache Spark

PostgreSQL

SAP HANA Cloud

Editions & Modules

No answers on this topic

Offerings

Pricing Offerings
Apache Spark	PostgreSQL	SAP HANA Cloud
Free Trial
No	No	Yes
Free/Freemium Version
No	No	No
Premium Consulting/Integration Services
No	No	No

Entry-level Setup Fee

No setup fee

Optional

Additional Details

—

Includes a one year free trial.

More Pricing Information

Pricing Info

Community Pulse
	Apache Spark	PostgreSQL	SAP HANA Cloud
Considered Multiple Products	Apache Spark Riyaz Khan Staff Engineer Chose Apache Spark Apache Spark is a fast-processing in-memory computing framework. It is 10 times faster than Apache Hadoop. Earlier we were using Apache Hadoop for processing data on the disk but now we are shifted to Apache Spark because of its in-memory computation capability. Also in SAP … Helpful? Verified User Executive Chose Apache Spark Databricks uses Spark as a foundation, and is also a great platform. It does bring several add-ons, which we did not feel needed by the time we evaluated - and haven't needed since then. One interesting plus in our opinion was the engineering support, which is great depending … Incentivized Helpful? SS Shiv Shivakumar Acquisitions Leader Chose Apache Spark We evaluated SAS alongside with Apache Spark but during the course of proof of concept found that Apache Spark was able to support the hadoop eco-system and hadoop file system much better. It was much faster at that time while having the ability to process data quickly for the … Incentivized Helpful?	PostgreSQL Verified User Employee Chose PostgreSQL MySQL is an Oracle product which has in itself some known issues due to that (support, contract terms). Based on my knowledge, PostgreSQL support everything that MySQL support (syntax wise) and it adds more improvements and syntaxes that make the life of database engineers and … Incentivized Helpful? Verified User Team Lead Chose PostgreSQL I found PostgreSQL to be better compared to MySQL. The community support is very good. Some features that I feel are not present in MySQL are: No referential integrity. No constraints (CHECK). Incentivized Helpful? Nitin Pasumarthy Research Assistant Chose PostgreSQL Compared to MySQL, it works well if you need to extend to your use case Compared to Spark, it works better w.r.t development time in a central database setting Like Redis, it cannot be used for caching and quick access of non-structured data Incentivized Helpful?	SAP HANA Cloud Verified User Engineer Chose SAP HANA Cloud We were using PostgreSQL prior to SAP HANA. The biggest difference that is noticed from an end-user standpoint is the speed with which database transactions take place. Because of the growing scale of our application, we really needed something faster. PostgreSQL just wasn’t … Incentivized Helpful? Verified User Administrator Chose SAP HANA Cloud speed wise its way faster than postgres and integration and UI for doing some db operations are out of box, there is no need of any client tools Incentivized Helpful? Verified User Engineer Chose SAP HANA Cloud Better tools and more integrated to BTP solutions. Incentivized Helpful? Verified User Engineer Chose SAP HANA Cloud On the BTP stack, SAP HANA Cloud has no alternative which can compete in terms of performance and integration to the overall platform Incentivized Helpful? Sreedhar Sree Sr. Automation QA in Machine Learning Chose SAP HANA Cloud As SAP HANA is an in-memory database, it can process data swiftly and can provide detailed analysis reports compared to other tools. Another advantage is it supports different data types, so if any application is looking for scalability, performance, security, and risk … Incentivized Helpful? Robert Forster SAP ERP Solution Architect Chose SAP HANA Cloud HANA has the best SAP Integration. Incentivized Helpful? SS Shiv Shivakumar Acquisitions Leader Chose SAP HANA Cloud We compared Microsoft BI with SAP HANA. The reasons to go with SAP HANA were - 1. ability to ingest data into HANA from a non SAP database 2. in-memory database resulting in faster real time analytics 3. ability to scale up 4. ability to replicate data real time 5. very solid … Incentivized Helpful?

Features

Apache Spark

PostgreSQL

SAP HANA Cloud

Relational Databases

Comparison of Relational Databases features of Product A and Product B
	Apache Spark - Ratings	PostgreSQL - Ratings	SAP HANA Cloud 7.8 27 Ratings 2% below category average
ACID compliance	00 Ratings	00 Ratings	8.420 Ratings
Database monitoring	00 Ratings	00 Ratings	7.726 Ratings
Database locking	00 Ratings	00 Ratings	7.922 Ratings
Encryption	00 Ratings	00 Ratings	7.623 Ratings
Disaster recovery	00 Ratings	00 Ratings	8.023 Ratings
Flexible deployment	00 Ratings	00 Ratings	7.525 Ratings
Multiple datatypes	00 Ratings	00 Ratings	7.625 Ratings

Best Alternatives
	Apache Spark	PostgreSQL	SAP HANA Cloud
Small Businesses	No answers on this topic	InfluxDB Score 8.8 out of 10	InterSystems IRIS Score 8.0 out of 10
Medium-sized Companies	Cloudera Manager Score 9.9 out of 10	SQLite Score 8.0 out of 10	InterSystems IRIS Score 8.0 out of 10
Enterprises	IBM Analytics Engine Score 7.2 out of 10	SQLite Score 8.0 out of 10	SAP IQ Score 10.0 out of 10
All Alternatives	View all alternatives	View all alternatives	View all alternatives

User Ratings
	Apache Spark	PostgreSQL	SAP HANA Cloud
Likelihood to Recommend	9.0 (24 ratings)	8.0 (55 ratings)	9.6 (308 ratings)
Likelihood to Renew	10.0 (1 ratings)	9.0 (1 ratings)	10.0 (11 ratings)
Usability	8.0 (4 ratings)	8.3 (9 ratings)	9.6 (29 ratings)
Availability	- (0 ratings)	9.0 (1 ratings)	3.6 (1 ratings)
Performance	- (0 ratings)	7.0 (1 ratings)	3.6 (1 ratings)
Support Rating	8.7 (4 ratings)	9.3 (7 ratings)	9.1 (251 ratings)
Implementation Rating	- (0 ratings)	9.0 (1 ratings)	9.1 (2 ratings)
Configurability	- (0 ratings)	- (0 ratings)	3.6 (1 ratings)
Ease of integration	- (0 ratings)	- (0 ratings)	4.5 (1 ratings)
Product Scalability	- (0 ratings)	8.0 (1 ratings)	4.5 (1 ratings)
Vendor post-sale	- (0 ratings)	- (0 ratings)	4.5 (1 ratings)
Vendor pre-sale	- (0 ratings)	- (0 ratings)	3.6 (1 ratings)

User Testimonials
	Apache Spark	PostgreSQL	SAP HANA Cloud
Likelihood to Recommend	Apache Well suited: To most of the local run of datasets and non-prod systems - scalability is not a problem at all. Including data from multiple types of data sources is an added advantage. MLlib is a decently nice built-in library that can be used for most of the ML tasks. Less appropriate: We had to work on a RecSys where the music dataset that we used was around 300+Gb in size. We faced memory-based issues. Few times we also got memory errors. Also the MLlib library does not have support for advanced analytics and deep-learning frameworks support. Understanding the internals of the working of Apache Spark for beginners is highly not possible. Incentivized Ananth Gouri Assistant Professor Read full review	PostgreSQL Global Development Group PostgreSQL is best used for structured data, and best when following relational database design principles. I would not use PostgreSQL for large unstructured data such as video, images, sound files, xml documents, web-pages, especially if these files have their own highly variable, internal structure. Incentivized Ron Ballard Consultant and author Read full review	SAP I think if you have a large organization, it's probably the product and the marketplace to go to. We're a large management consulting firm operating in four to seven countries. And generally speaking, I think that's the size and the scope where it scales best. I can't speak to smaller companies, but I can't see smaller companies leveraging the benefits as much as a larger organization can. Incentivized Verified User Anonymous Read full review
Pros	Apache Rich APIs for data transformation making for very each to transform and prepare data in a distributed environment without worrying about memory issues Faster in execution times compare to Hadoop and PIG Latin Easy SQL interface to the same data set for people who are comfortable to explore data in a declarative manner Interoperability between SQL and Scala / Python style of munging data Incentivized Nitin Pasumarthy Software Engineer Read full review	PostgreSQL Global Development Group It works well with external data sources and runs on platforms with stable performance. Clients can rest assured that their personal information will be safe and secure. Many forums discuss setup and usage, and most are free. Adding tooling applications to a computer is unlimited. PostgreSQL runs on many OS platforms and supports ANSI SQL, stored procedures, and triggers. Aurpa Fiza Software Application Developer Read full review	SAP Real-time reporting and analytics on data: because of its in-memory architecture, it is perfect for businesses that need to make quick decisions based on current information. Managing workload with complex data: it can handle a vast range of data types, including relational, documental, geospatial, graph, vector, and time series data. Developing and deploying intelligent data applications: it provides various tools for such applications and can be used for machine learning and artificial intelligence to automate tasks, gain insights from data, and make predictions. SP sonam pahwa SAP consultant Read full review
Cons	Apache Memory management. Very weak on that. PySpark not as robust as scala with spark. spark master HA is needed. Not as HA as it should be. Locality should not be a necessity, but does help improvement. But would prefer no locality Incentivized Anson Abraham Data Czar Read full review	PostgreSQL Global Development Group Clearer indications on what is the query plan, to optimize the query More out of the box, Postgres specific, SQL functions It would be nice to have a more visual aid of the relationship between all tables, but possibly this depend more on the UI used Incentivized Verified User Anonymous Read full review	SAP Requires higher processing power, otherwise it won't fly. How ever computing costs are lower. Incase you are migrating to cloud please do not select the highest config available in that series . Upgrading it later against a reserved instance can cost you dearly with a series change Lack of clarity on licensing is one major challenge Unless S/4 with additional features are enabled mere migration HANA DB is not a rewarding journey. Power is in S/4 Incentivized Verified User Anonymous Read full review
Likelihood to Renew	Apache Capacity of computing data in cluster and fast speed. Steven Li Senior Software Developer (Consultant) Read full review	PostgreSQL Global Development Group As a needed software for day to day development activities Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	SAP We would rate our likelihood of renewing at 9/10. SAP HANA Cloud has proven to be a highly reliable and scalable data platform that consistently delivers strong performance. Its seamless integration with our overall SAP landscape, combined with improved analytics and real-time data capabilities, makes it a core part of our long-term technology strategy. Incentivized Verified User Anonymous Read full review
Usability	Apache If the team looking to use Apache Spark is not used to debug and tweak settings for jobs to ensure maximum optimizations, it can be frustrating. However, the documentation and the support of the community on the internet can help resolve most issues. Moreover, it is highly configurable and it integrates with different tools (eg: it can be used by dbt core), which increase the scenarios where it can be used Incentivized Verified User Anonymous Read full review	PostgreSQL Global Development Group Postgresql is the best tool out there for relational data so I have to give it a high rating when it comes to analytics, data availability and consistency, so on and so forth. SQL is also a relatively consistent language so when it comes to building new tables and loading data in from the OLTP database, there are enough tools where we can perform ETL on a scalable basis. Incentivized Verified User Anonymous Read full review	SAP It is very useful solution which provides you speedier data processing, real-time analytics. It helps you manage diverse data types. It also offers you excellent disaster management. It has user friendly interface which helps you navigate system and transactions easily and perform task smoothly. Incentivized GM Gopal Mishra IT Executive Read full review
Reliability and Availability	Apache No answers on this topic	PostgreSQL Global Development Group PostgreSQL's availability is top notch. Apart from connection time-out for an idle user, the database is super reliable. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	SAP so far, we didn't get any outage Incentivized Verified User Anonymous Read full review
Performance	Apache No answers on this topic	PostgreSQL Global Development Group The data queries are relatively quick for a small to medium sized table. With complex joins, and a wide and deep table however, the performance of the query has room for improvement. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	SAP so far good Incentivized Verified User Anonymous Read full review
Support Rating	Apache 1. It integrates very well with scala or python. 2. It's very easy to understand SQL interoperability. 3. Apache is way faster than the other competitive technologies. 4. The support from the Apache community is very huge for Spark. 5. Execution times are faster as compared to others. 6. There are a large number of forums available for Apache Spark. 7. The code availability for Apache Spark is simpler and easy to gain access to. 8. Many organizations use Apache Spark, so many solutions are available for existing applications. YM Yogesh Mhasde Technical Manager Read full review	PostgreSQL Global Development Group There are several companies that you can contract for technical support, like EnterpriseDB or Percona, both first level in expertise and commitment to the software. But we do not have contracts with them, we have done all the way from googling to forums, and never have a problem that we cannot resolve or pass around. And for dozens of projects and more than 15 years now. Incentivized Javier Blanque Technology Risk and Information Assets Manager Read full review	SAP However, I am not the right person to answer this as we have another department to handle support and contact the service provider for any support required. Although i will say that they are the quick respondent and knows how to handle querry of the customers and provide quick and better support. Incentivized Verified User Anonymous Read full review
Online Training	Apache No answers on this topic	PostgreSQL Global Development Group The online training is request based. Had there been recorded videos available online for potential users to benefit from, I could have rated it higher. The online documentation however is very helpful. The online documentation PDF is downloadable and allows users to pace their own learning. With examples and code snippets, the documentation is great starting point. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	SAP No answers on this topic
Implementation Rating	Apache No answers on this topic	PostgreSQL Global Development Group The online documentation of the PostgreSQL product is elaborate and takes users step by step. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	SAP Professional GIS people are some of the most risk-averse there are, and it's difficult to get them to move to HANA in one step. Start with small projects building to 80% use of HANA spatial over time. Incentivized Verified User Anonymous Read full review
Alternatives Considered	Apache Spark in comparison to similar technologies ends up being a one stop shop. You can achieve so much with this one framework instead of having to stitch and weave multiple technologies from the Hadoop stack, all while getting incredibility performance, minimal boilerplate, and getting the ability to write your application in the language of your choosing. Incentivized Verified User Anonymous Read full review	PostgreSQL Global Development Group Although the competition between the different databases is increasingly aggressive in the sense that they provide many improvements, new functionalities, compatibility with complementary components or environments, in some cases it requires that it be followed within the same family of applications that performs the company that develops it and that is not all bad, but being able to adapt or configure different programs, applications or other environments developed by third parties apart is what gives PostgreSQL a certain advantage and this diversification in the components that can be joined with it, is the reason why it is a great option to choose. Incentivized Moris Mendez Ing. de Sistemas Informaticos Read full review	SAP I have deep knowledge of other disk based DBMSs. They are venerable technology, but the attempts to extend them to current architectures belie the fact they are built on 40 year old technology. There are some good columnar in-memory databases but they lack the completeness of capability present in the HANA platform. Incentivized TT Tom Turchioe AP Ecosystem Technology Lead - Global Read full review
Contract Terms and Pricing Model	Apache No answers on this topic	PostgreSQL Global Development Group No answers on this topic	SAP I don't have visibility in licensing Incentivized Verified User Anonymous Read full review
Scalability	Apache No answers on this topic	PostgreSQL Global Development Group The DB is reliable, scalable, easy to use and resolves most DB needs Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	SAP Limitation of training deliverable by organization Incentivized SA Shoaib Ansari DEVELOPER Read full review
Professional Services	Apache No answers on this topic	PostgreSQL Global Development Group No answers on this topic	SAP We are still in process for the first applciaiton Incentivized Verified User Anonymous Read full review
Return on Investment	Apache Business leaders are able to take data driven decisions Business users are able access to data in near real time now . Before using spark, they had to wait for at least 24 hours for data to be available Business is able come up with new product ideas Incentivized Surendranatha Reddy Chappidi Senior Data Engineer Read full review	PostgreSQL Global Development Group Easy to administer so our DevOps team has only ever used minimal time to setup, tune, and maintain. Easy to interface with so our Engineering team has only ever used minimal time to query or modify the database. Getting the data is straightforward, what we do with it is the bigger concern. It's free. You can't beat that. Incentivized Don Burks Technical Lead Read full review	SAP ROI has always been high in terms of the functionality that it offers and the security features it comes with. Managing large volumes of data in real-time is not an easy task, but it does it pretty well with faster data processing. Incentivized Nishant Kumar Lead Engineer Read full review
ScreenShots