Apache Sqoop vs. IBM Db2 Big SQL

Apache Sqoop

4 Reviews and Ratings

IBM Db2 Big SQL

IBM Db2 Big SQL

16 Reviews and Ratings

Overview
Product	Rating	Most Used By	Product Summary	Starting Price
Apache Sqoop	Score 8.8 out of 10	N/A	Apache Sqoop is a tool for use with Hadoop, used to transfer data between Apache Hadoop and other, structured data stores.	N/A
Db2 Big SQL	Score 8.7 out of 10	N/A	IBM offers Db2 Big SQL, an enterprise grade hybrid ANSI-compliant SQL on Hadoop engine, delivering massively parallel processing (MPP) and advanced data query. Big SQL offers a single database connection or query for disparate sources such as HDFS, RDMS, NoSQL databases, object stores and WebHDFS.	N/A

Pricing

Apache Sqoop

IBM Db2 Big SQL

Editions & Modules

No answers on this topic

No answers on this topic

Offerings

Pricing Offerings
Apache Sqoop	Db2 Big SQL
Free Trial
No	No
Free/Freemium Version
No	No
Premium Consulting/Integration Services
No	No

Entry-level Setup Fee

No setup fee

No setup fee

Additional Details

—

—

More Pricing Information

Community Pulse
	Apache Sqoop	IBM Db2 Big SQL
Top Pros	Pro Supports multiple Pro External data Pro Data sources	Pro Data storage Pro Data manipulation Pro High performance
Top Cons	Minus Multiple tables Minus Time consuming	Minus Ease of implementation

Best Alternatives
	Apache Sqoop	IBM Db2 Big SQL
Small Businesses	No answers on this topic	No answers on this topic
Medium-sized Companies	Cloudera Manager Score 9.7 out of 10	Cloudera Manager Score 9.7 out of 10
Enterprises	IBM Analytics Engine Score 8.8 out of 10	IBM Analytics Engine Score 8.8 out of 10
All Alternatives	View all alternatives	View all alternatives

User Ratings
	Apache Sqoop	IBM Db2 Big SQL
Likelihood to Recommend	9.0 (1 ratings)	9.0 (2 ratings)
Usability	- (0 ratings)	8.0 (1 ratings)
Support Rating	- (0 ratings)	8.8 (2 ratings)

User Testimonials
	Apache Sqoop	IBM Db2 Big SQL
Likelihood to Recommend	Apache Sqoop is great for sending data between a JDBC compliant database and a Hadoop environment. Sqoop is built for those who need a few simple CLI options to import a selection of database tables into Hadoop, do large dataset analysis that could not commonly be done with that database system due to resource constraints, then export the results back into that database (or another). Sqoop falls short when there needs to be some extra, customized processing between database extract, and Hadoop loading, in which case Apache Spark's JDBC utilities might be preferred Incentivized Jordan Moore Consultant Read full review	IBM My recommendation obviously would depend on the application. But I think given the right requirements, IBM DB2 Big SQL is definitely a contender for a database platform. Especially when disparate data and multiple data stores are involved. I like the fact I can use the product to federate my data and make it look like it's all in one place. The engine is high performance and if you desire to use Hadoop, this could be your platform. Incentivized Gene Baker Vice President, Chief Architect, Development Manager and Software Engineer Read full review
Pros	Apache Provides generalized JDBC extensions to migrate data between most database systems Generates Java classes upon reading database records for use in other code utilizing Hadoop's client libraries Allows for both import and export features Incentivized Jordan Moore Consultant Read full review	IBM data storage data manipulation data definitions data reliability Incentivized JS John Spies Database Administrator Read full review
Cons	Apache Sqoop2 development seems to have stalled. I have set it up outside of a Cloudera CDH installation, and I actually prefer it's "Sqoop Server" model better than just the CLI client version that is Sqoop1. This works especially well in a microservices environment, where there would be only one place to maintain the JDBC drivers to use for Sqoop. Incentivized Jordan Moore Consultant Read full review	IBM Cloud readiness. Ease of implementation. Incentivized Gene Baker Vice President, Chief Architect, Development Manager and Software Engineer Read full review
Usability	Apache No answers on this topic	IBM IBM DB2 is a solid service but hasn't seen much innovation over the past decade. It gets the job done and supports our IT operations across digital so it is fair. Incentivized JS John Spies Database Administrator Read full review
Support Rating	Apache No answers on this topic	IBM IBM did a good job of supporting us during our evaluation and proof of concept. They were able to provide all necessary guidance, answer questions, help us architect it, etc. We were pleased with the support provided by the vendor. I will caveat and say this support was all before the sale, however, we have a ton of IBM products and they provide the same high level of support for all of them. I didn't see this being any different. I give IBM support two thumbs up! Incentivized Gene Baker Vice President, Chief Architect, Development Manager and Software Engineer Read full review
Alternatives Considered	Apache Sqoop comes preinstalled on the major Hadoop vendor distributions as the recommended product to import data from relational databases. The ability to extend it with additional JDBC drivers makes it very flexible for the environment it is installed within. Spark also has a useful JDBC reader, and can manipulate data in more ways than Sqoop, and also upload to many other systems than just Hadoop. Kafka Connect JDBC is more for streaming database updates using tools such as Oracle GoldenGate or Debezium. Streamsets and Apache NiFi both provide a more "flow based programming" approach to graphically laying out connectors between various systems, including JDBC and Hadoop. Incentivized Jordan Moore Consultant Read full review	IBM MS SQL Server was ruled out given we didn't feel we could collapse environments. We thought of MS-SQL as more of a one for one replacement for Sybase ASE, i.e., server for server. SAP HANA was evaluated and given a big thumbs up but was rejected because the SQL would have to be rewritten at the time (now they have an accelerator so you don't have to). Also, there was a very low adoption rate within the enterprise. IBM DB2 Big SQL was not selected even though technically it achieved high scores, because we could not find readily available talent and low adoption rate within the enterprise (basically no adoption at the time). We ended up selecting Exadata because of the high adoption rate within the enterprise even though technically HANA and Big SQL were superior in our evaluations. Incentivized Gene Baker Vice President, Chief Architect, Development Manager and Software Engineer Read full review
Return on Investment	Apache When combined with Cloudera's HUE, it can enable non-technical users to easily import relational data into Hadoop. Being able to manipulate large datasets in Hadoop, and them load them into a type of "materialized view" in an external database system has yielded great insights into the Hadoop datalake without continuously running large batch jobs. Sqoop isn't very user-friendly for those uncomfortable with a CLI. Incentivized Jordan Moore Consultant Read full review	IBM better data visibility solid reliability for mission critical data Incentivized JS John Spies Database Administrator Read full review
ScreenShots