Apache Sqoop vs. PostgreSQL vs. Presto

Apache Sqoop

Apache Sqoop

4 Reviews and Ratings

PostgreSQL

PostgreSQL

354 Reviews and Ratings

Presto

Presto

19 Reviews and Ratings

Overview
Product	Rating	Most Used By	Product Summary	Starting Price
Apache Sqoop	Score 8.8 out of 10	N/A	Apache Sqoop is a tool for use with Hadoop, used to transfer data between Apache Hadoop and other, structured data stores.	N/A
PostgreSQL	Score 8.7 out of 10	N/A	PostgreSQL (alternately Postgres) is a free and open source object-relational database system boasting over 30 years of active development, reliability, feature robustness, and performance. It supports SQL and is designed to support various workloads flexibly.	N/A
Presto	Score 10.0 out of 10	N/A	Presto is an open source SQL query engine designed to run queries on data stored in Hadoop or in traditional databases. Teradata supported development of Presto followed the acquisition of Hadapt and Revelytix.	N/A

Pricing

Apache Sqoop

PostgreSQL

Presto

Editions & Modules

No answers on this topic

No answers on this topic

No answers on this topic

Offerings

Pricing Offerings
Apache Sqoop	PostgreSQL	Presto
Free Trial
No	No	No
Free/Freemium Version
No	No	No
Premium Consulting/Integration Services
No	No	No

Entry-level Setup Fee

No setup fee

No setup fee

No setup fee

Additional Details

—

—

—

More Pricing Information

Community Pulse
	Apache Sqoop	PostgreSQL	Presto

Best Alternatives
	Apache Sqoop	PostgreSQL	Presto
Small Businesses	No answers on this topic	InfluxDB Score 8.8 out of 10	InterSystems IRIS Score 8.0 out of 10
Medium-sized Companies	Cloudera Manager Score 9.9 out of 10	SQLite Score 8.0 out of 10	InterSystems IRIS Score 8.0 out of 10
Enterprises	IBM Analytics Engine Score 7.2 out of 10	SQLite Score 8.0 out of 10	SAP IQ Score 10.0 out of 10
All Alternatives	View all alternatives	View all alternatives	View all alternatives

User Ratings
	Apache Sqoop	PostgreSQL	Presto
Likelihood to Recommend	9.0 (1 ratings)	8.0 (55 ratings)	7.8 (2 ratings)
Likelihood to Renew	- (0 ratings)	9.0 (1 ratings)	- (0 ratings)
Usability	- (0 ratings)	8.3 (9 ratings)	- (0 ratings)
Availability	- (0 ratings)	9.0 (1 ratings)	- (0 ratings)
Performance	- (0 ratings)	7.0 (1 ratings)	- (0 ratings)
Support Rating	- (0 ratings)	9.3 (7 ratings)	- (0 ratings)
Implementation Rating	- (0 ratings)	9.0 (1 ratings)	- (0 ratings)
Product Scalability	- (0 ratings)	8.0 (1 ratings)	- (0 ratings)

User Testimonials
	Apache Sqoop	PostgreSQL	Presto
Likelihood to Recommend	Apache Sqoop is great for sending data between a JDBC compliant database and a Hadoop environment. Sqoop is built for those who need a few simple CLI options to import a selection of database tables into Hadoop, do large dataset analysis that could not commonly be done with that database system due to resource constraints, then export the results back into that database (or another). Sqoop falls short when there needs to be some extra, customized processing between database extract, and Hadoop loading, in which case Apache Spark's JDBC utilities might be preferred Incentivized Jordan Moore Consultant Read full review	PostgreSQL Global Development Group PostgreSQL is best used for structured data, and best when following relational database design principles. I would not use PostgreSQL for large unstructured data such as video, images, sound files, xml documents, web-pages, especially if these files have their own highly variable, internal structure. Incentivized Ron Ballard Consultant and author Read full review	Open Source Presto is for interactive simple queries, where Hive is for reliable processing. If you have a fact-dim join, presto is great..however for fact-fact joins presto is not the solution.. Presto is a great replacement for proprietary technology like Vertica Incentivized Praveen Murugesan Engineering Manager - Ride Experience Read full review
Pros	Apache Provides generalized JDBC extensions to migrate data between most database systems Generates Java classes upon reading database records for use in other code utilizing Hadoop's client libraries Allows for both import and export features Incentivized Jordan Moore Consultant Read full review	PostgreSQL Global Development Group It works well with external data sources and runs on platforms with stable performance. Clients can rest assured that their personal information will be safe and secure. Many forums discuss setup and usage, and most are free. Adding tooling applications to a computer is unlimited. PostgreSQL runs on many OS platforms and supports ANSI SQL, stored procedures, and triggers. Aurpa Fiza Software Application Developer Read full review	Open Source Linking, embedding links and adding images is easy enough. Once you have become familiar with the interface, Presto becomes very quick & easy to use (but, you have to practice & repeat to know what you are doing - it is not as intuitive as one would hope). Organizing & design is fairly simple with click & drag parameters. Incentivized Corinne Nacin-Martinez Consumer Sales & Sevice Manager Read full review
Cons	Apache Sqoop2 development seems to have stalled. I have set it up outside of a Cloudera CDH installation, and I actually prefer it's "Sqoop Server" model better than just the CLI client version that is Sqoop1. This works especially well in a microservices environment, where there would be only one place to maintain the JDBC drivers to use for Sqoop. Incentivized Jordan Moore Consultant Read full review	PostgreSQL Global Development Group Clearer indications on what is the query plan, to optimize the query More out of the box, Postgres specific, SQL functions It would be nice to have a more visual aid of the relationship between all tables, but possibly this depend more on the UI used Incentivized Verified User Anonymous Read full review	Open Source Presto was not designed for large fact fact joins. This is by design as presto does not leverage disk and used memory for processing which in turn makes it fast.. However, this is a tradeoff..in an ideal world, people would like to use one system for all their use cases, and presto should get exhaustive by solving this problem. Resource allocation is not similar to YARN and presto has a priority queue based query resource allocation..so a query that takes long takes longer...this might be alleviated by giving some more control back to the user to define priority/override. UDF Support is not available in presto. You will have to write your own functions..while this is good for performance, it comes at a huge overhead of building exclusively for presto and not being interoperable with other systems like Hive, SparkSQL etc. Incentivized Praveen Murugesan Engineering Manager - Ride Experience Read full review
Likelihood to Renew	Apache No answers on this topic	PostgreSQL Global Development Group As a needed software for day to day development activities Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	Open Source No answers on this topic
Usability	Apache No answers on this topic	PostgreSQL Global Development Group Postgresql is the best tool out there for relational data so I have to give it a high rating when it comes to analytics, data availability and consistency, so on and so forth. SQL is also a relatively consistent language so when it comes to building new tables and loading data in from the OLTP database, there are enough tools where we can perform ETL on a scalable basis. Incentivized Verified User Anonymous Read full review	Open Source No answers on this topic
Reliability and Availability	Apache No answers on this topic	PostgreSQL Global Development Group PostgreSQL's availability is top notch. Apart from connection time-out for an idle user, the database is super reliable. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	Open Source No answers on this topic
Performance	Apache No answers on this topic	PostgreSQL Global Development Group The data queries are relatively quick for a small to medium sized table. With complex joins, and a wide and deep table however, the performance of the query has room for improvement. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	Open Source No answers on this topic
Support Rating	Apache No answers on this topic	PostgreSQL Global Development Group There are several companies that you can contract for technical support, like EnterpriseDB or Percona, both first level in expertise and commitment to the software. But we do not have contracts with them, we have done all the way from googling to forums, and never have a problem that we cannot resolve or pass around. And for dozens of projects and more than 15 years now. Incentivized Javier Blanque Technology Risk and Information Assets Manager Read full review	Open Source No answers on this topic
Online Training	Apache No answers on this topic	PostgreSQL Global Development Group The online training is request based. Had there been recorded videos available online for potential users to benefit from, I could have rated it higher. The online documentation however is very helpful. The online documentation PDF is downloadable and allows users to pace their own learning. With examples and code snippets, the documentation is great starting point. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	Open Source No answers on this topic
Implementation Rating	Apache No answers on this topic	PostgreSQL Global Development Group The online documentation of the PostgreSQL product is elaborate and takes users step by step. Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	Open Source No answers on this topic
Alternatives Considered	Apache Sqoop comes preinstalled on the major Hadoop vendor distributions as the recommended product to import data from relational databases. The ability to extend it with additional JDBC drivers makes it very flexible for the environment it is installed within. Spark also has a useful JDBC reader, and can manipulate data in more ways than Sqoop, and also upload to many other systems than just Hadoop. Kafka Connect JDBC is more for streaming database updates using tools such as Oracle GoldenGate or Debezium. Streamsets and Apache NiFi both provide a more "flow based programming" approach to graphically laying out connectors between various systems, including JDBC and Hadoop. Incentivized Jordan Moore Consultant Read full review	PostgreSQL Global Development Group Although the competition between the different databases is increasingly aggressive in the sense that they provide many improvements, new functionalities, compatibility with complementary components or environments, in some cases it requires that it be followed within the same family of applications that performs the company that develops it and that is not all bad, but being able to adapt or configure different programs, applications or other environments developed by third parties apart is what gives PostgreSQL a certain advantage and this diversification in the components that can be joined with it, is the reason why it is a great option to choose. Incentivized Moris Mendez Ing. de Sistemas Informaticos Read full review	Open Source Presto is good for a templated design appeal. You cannot be too creative via this interface - but, the layout and options make the finalized visual product appealing to customers. The other design products I use are for different purposes and not really comparable to Presto. Incentivized Corinne Nacin-Martinez Consumer Sales & Sevice Manager Read full review
Scalability	Apache No answers on this topic	PostgreSQL Global Development Group The DB is reliable, scalable, easy to use and resolves most DB needs Incentivized Ojoswi Basu Sr. Tableau Solution Consultant Read full review	Open Source No answers on this topic
Return on Investment	Apache When combined with Cloudera's HUE, it can enable non-technical users to easily import relational data into Hadoop. Being able to manipulate large datasets in Hadoop, and them load them into a type of "materialized view" in an external database system has yielded great insights into the Hadoop datalake without continuously running large batch jobs. Sqoop isn't very user-friendly for those uncomfortable with a CLI. Incentivized Jordan Moore Consultant Read full review	PostgreSQL Global Development Group Easy to administer so our DevOps team has only ever used minimal time to setup, tune, and maintain. Easy to interface with so our Engineering team has only ever used minimal time to query or modify the database. Getting the data is straightforward, what we do with it is the bigger concern. It's free. You can't beat that. Incentivized Don Burks Technical Lead Read full review	Open Source Presto has helped scale Uber's interactive data needs. We have migrated a lot out of proprietary tech like Vertica. Presto has helped build data driven applications on its stack than maintain a separate online/offline stack. Presto has helped us build data exploration tools by leveraging it's power of interactive and is immensely valuable for data scientists. Incentivized Praveen Murugesan Engineering Manager - Ride Experience Read full review
ScreenShots