TrustRadius: an HG Insights company

Apache Sqoop (discontinued) vs. IBM Confluent

Save this comparison

Save this comparison

Add Product

Recommended Comparisons

    Overview
    ProductRatingMost Used ByProduct SummaryStarting Price

    Apache Sqoop (discontinued)

    Score8.8 out of 10
    N/AApache Sqoop is a discontinued open-source command-line tool for bulk data transfer between Apache Hadoop and structured data stores. It was commonly used to import relational database tables or mainframe datasets into HDFS for processing with Hadoop tools, then export processed data back to a relational database.N/A

    IBM Confluent

    Score9.1 out of 10
    N/AIBM Confluent helps enterprises stream, connect, process, govern and serve real-time data across hybrid environments.

    $385

    per month

    Pricing
    Apache Sqoop (discontinued)IBM Confluent
    Editions & Modules
    No answers on this topic
    Basic
    $0
    Standard
    Starting at ~$385
    per month
    Enterprise
    Starting at ~$1,150
    per month
    Offerings
    Pricing Offerings
    Apache Sqoop (discontinued)IBM Confluent
    Free Trial
    NoNo
    Free/Freemium Version
    NoYes
    Premium Consulting/Integration Services
    NoNo
    Entry-level Setup FeeNo setup feeNo setup fee
    Additional DetailsConfluent monthly bills are based upon resource consumption, i.e., you are only charged for the resources you use when you actually use them: Stream: Kafka clusters are billed for eCKUs/CKUs ($/hour), networking ($/GB), and storage ($/GB-hour). Connect: Use of connectors is billed based on throughput ($/GB) and a task base price ($/task/hour). Process: Use of stream processing with Confluent Cloud for Apache Flink is calculated based on CFUs ($/minute). Govern: Use of Stream Governance is billed based on environment ($/hour). Confluent storage and throughput is calculated in binary gigabytes (GB), where 1 GB is 2^30 bytes. This unit of measurement is also known as a gibibyte (GiB). Please also note that all prices are stated in United States Dollars unless specifically stated otherwise. All billing computations are conducted in Coordinated Universal Time (UTC).
    More Pricing Information
    Best Alternatives
    Apache Sqoop (discontinued)IBM Confluent
    Small Businesses
    No answers on this topic
    Amazon SNS
    Score8.7 out of 10
    Medium-sized Companies
    Apache Spark
    Score8.8 out of 10
    Amazon SNS
    Score8.7 out of 10
    Enterprises
    Apache Spark
    Score8.8 out of 10
    Google Cloud Pub/Sub
    Score8.8 out of 10
    All AlternativesView all alternativesView all alternatives
    User Ratings
    Apache Sqoop (discontinued)IBM Confluent
    Likelihood to Recommend
    9.0
    (1 ratings)
    10.0
    (2 ratings)
    Support Rating
    -
    (0 ratings)
    10.0
    (1 ratings)
    User Testimonials
    Apache Sqoop (discontinued)IBM Confluent
    Likelihood to Recommend
    Apache
    Sqoop is great for sending data between a JDBC compliant database and a Hadoop environment. Sqoop is built for those who need a few simple CLI options to import a selection of database tables into Hadoop, do large dataset analysis that could not commonly be done with that database system due to resource constraints, then export the results back into that database (or another). Sqoop falls short when there needs to be some extra, customized processing between database extract, and Hadoop loading, in which case Apache Spark's JDBC utilities might be preferred
    Incentivized
    Read full review
    IBM
    If you have a need to stream data, real time or segmented structured data then Confluent is a great platform to do so with. You won't run into packet transfer size limitations that other platforms have. Flexibility in on-prem, cloud, and managed cloud offerings makes it very flexible no matter how you choose to implement.
    Incentivized
    Read full review
    Pros
    Apache
    • Provides generalized JDBC extensions to migrate data between most database systems
    • Generates Java classes upon reading database records for use in other code utilizing Hadoop's client libraries
    • Allows for both import and export features
    Incentivized
    Read full review
    IBM
    • Products work great.
    • Training is available.
    • Customer support is good.
    Incentivized
    Read full review
    Cons
    Apache
    • Sqoop2 development seems to have stalled. I have set it up outside of a Cloudera CDH installation, and I actually prefer it's "Sqoop Server" model better than just the CLI client version that is Sqoop1. This works especially well in a microservices environment, where there would be only one place to maintain the JDBC drivers to use for Sqoop.
    Incentivized
    Read full review
    IBM
    • Cloud based Azure platform features for Confluent lacks behind AWS And GCP
    Incentivized
    Read full review
    Support Rating
    Apache
    No answers on this topic
    IBM
    The support from the Confluent platform is great and satisfying. We have been working with Confluent for more than a year now. They sent out resident architects to help us set up Confluent cluster on our cloud and help us troubleshoot problems we have encountered. Overall, it has been a great experience working with the Confluent Platform.
    Incentivized
    Read full review
    Alternatives Considered
    Apache
    • Sqoop comes preinstalled on the major Hadoop vendor distributions as the recommended product to import data from relational databases. The ability to extend it with additional JDBC drivers makes it very flexible for the environment it is installed within.
    • Spark also has a useful JDBC reader, and can manipulate data in more ways than Sqoop, and also upload to many other systems than just Hadoop.
    • Kafka Connect JDBC is more for streaming database updates using tools such as Oracle GoldenGate or Debezium.
    • Streamsets and Apache NiFi both provide a more "flow based programming" approach to graphically laying out connectors between various systems, including JDBC and Hadoop.
    Incentivized
    Read full review
    IBM
    For our use case it was very important that the technology we were working with fit into our Azure architecture, and met our data processing size requirements to stream data within certain SLAs. Confluent more than met our performance requirements and compared to the others scale options and cost to run it was more than financially viable as a platform solution to our global operations.
    Incentivized
    Read full review
    Return on Investment
    Apache
    • When combined with Cloudera's HUE, it can enable non-technical users to easily import relational data into Hadoop.
    • Being able to manipulate large datasets in Hadoop, and them load them into a type of "materialized view" in an external database system has yielded great insights into the Hadoop datalake without continuously running large batch jobs.
    • Sqoop isn't very user-friendly for those uncomfortable with a CLI.
    Incentivized
    Read full review
    IBM
    • It enables us to develop event driven application.
    • It increases our ability to handle streaming data.
    • It reduces latency of communication.
    Incentivized
    Read full review
    ScreenShots