What users are saying about
127 Ratings
22 Ratings
127 Ratings
<a href='https://www.trustradius.com/static/about-trustradius-scoring' target='_blank' rel='nofollow noopener noreferrer'>trScore algorithm: Learn more.</a>
Score 8.7 out of 100
22 Ratings
<a href='https://www.trustradius.com/static/about-trustradius-scoring' target='_blank' rel='nofollow noopener noreferrer'>trScore algorithm: Learn more.</a>
Score 3.2 out of 100

Likelihood to Recommend

Apache Spark

The software appears to run more efficiently than other big data tools, such as Hadoop. Given that, Apache Spark is well-suited for querying and trying to make sense of very, very large data sets. The software offers many advanced machine learning and econometrics tools, although these tools are used only partially because very large data sets require too much time when the data sets get too large. The software is not well-suited for projects that are not big data in size. The graphics and analytical output are subpar compared to other tools.
Thomas Young | TrustRadius Reviewer

Informatica MDM

If your environment is clean and well organized, but at the same time has many application domains with their own data sources, Informatica MDM can be a really good factor in maintaining "single points of truth" for all this data. However, the more application domains you have, the less "clean" your environment becomes. If your application domain landscape consists of multiple technologies (C#, Oracle, JAVA, web-based, windows services, console apps, third-party tools, etc), your environment becomes a real nightmare to maintain unless you implement a service-oriented approach. And this is where Informatica MDM fails completely since it promotes a "point to point" scenario. At least this is my experience. It could be that Informatica MDM supports a service-oriented approach, but I have not seen this. I could be that the developers in my organization who have expert Informatica MDM knowledge are just more "point to point" oriented. But even if that is the case, it's a valid argument against Informatica MDM, since it's already hard enough to find developers who are dedicated to this product, it becomes impossible to find SOA oriented developers.
Anonymous | TrustRadius Reviewer

Pros

Apache Spark

  • Rich APIs for data transformation making for very each to transform and prepare data in a distributed environment without worrying about memory issues
  • Faster in execution times compare to Hadoop and PIG Latin
  • Easy SQL interface to the same data set for people who are comfortable to explore data in a declarative manner
  • Interoperability between SQL and Scala / Python style of munging data
Nitin Pasumarthy | TrustRadius Reviewer

Informatica MDM

  • Gather data.
  • Manage data.
  • Present data to the user.
  • Correlate data.
  • Data manipulation.
Anonymous | TrustRadius Reviewer

Cons

Apache Spark

  • Memory management. Very weak on that.
  • PySpark not as robust as scala with spark.
  • spark master HA is needed. Not as HA as it should be.
  • Locality should not be a necessity, but does help improvement. But would prefer no locality
Anson Abraham | TrustRadius Reviewer

Informatica MDM

  • Mapping within the tool can be difficult. But that is slated for upgrade with the next version.
  • Many detailed screens for the developer interface. This makes it hard to find options sometimes.
  • Address validation sometimes leads to incorrect results. This is my biggest issue with the product.
Brian Randolph | TrustRadius Reviewer

Usability

Apache Spark

Apache Spark 8.7
Based on 3 answers
Apache integrates with multiple big data frameworks. It does not exert too much load on the disks. Moreover, it is easy to program and use. It reduces the headache of using different applications separately through its high-level APIs. Big data processing has never been as easy as it is with Apache Spark.
Partha Protim Pegu | TrustRadius Reviewer

Informatica MDM

No score
No answers yet
No answers on this topic

Support Rating

Apache Spark

Apache Spark 8.2
Based on 6 answers
1. It integrates very well with scala or python.2. It's very easy to understand SQL interoperability.3. Apache is way faster than the other competitive technologies.4. The support from the Apache community is very huge for Spark.5. Execution times are faster as compared to others.6. There are a large number of forums available for Apache Spark.7. The code availability for Apache Spark is simpler and easy to gain access to.8. Many organizations use Apache Spark, so many solutions are available for existing applications.
Yogesh Mhasde | TrustRadius Reviewer

Informatica MDM

Informatica MDM 8.0
Based on 1 answer
I'm not sure since I never used support. My colleagues never had any issues with it, therefore my rating would be an 8 with a certain range of uncertainty.
Anonymous | TrustRadius Reviewer

Alternatives Considered

Apache Spark

Spark in comparison to similar technologies ends up being a one stop shop. You can achieve so much with this one framework instead of having to stitch and weave multiple technologies from the Hadoop stack, all while getting incredibility performance, minimal boilerplate, and getting the ability to write your application in the language of your choosing.
Anonymous | TrustRadius Reviewer

Informatica MDM

Less expensive then the competitors, modules are self-sustained and able to function directly. It involves less coding, and a lot of things can be done in house!
Anonymous | TrustRadius Reviewer

Return on Investment

Apache Spark

  • It has had a very positive impact, as it helps reduce the data processing time and thus helps us achieve our goals much faster.
  • Being easy to use, it allows us to adapt to the tool much faster than with others, which in turn allows us to access various data sources such as Hadoop, Apache Mesos, Kubernetes, independently or in the cloud. This makes it very useful.
  • It was very easy for me to use Apache Spark and learn it since I come from a background of Java and SQL, and it shares those basic principles and uses a very similar logic.
Carla Borges | TrustRadius Reviewer

Informatica MDM

  • Informatica MDM allowed for a faster standup of our Salesforce environment.
  • Informatica MDM provided a means to keep our multiple source systems synchronized.
  • Informatica MDM can be pricey but is definitely a leader in the space.
Brian Randolph | TrustRadius Reviewer

Pricing Details

Apache Spark

General

Free Trial
Free/Freemium Version
Premium Consulting/Integration Services
Entry-level set up fee?
No

Informatica MDM

General

Free Trial
Free/Freemium Version
Premium Consulting/Integration Services
Entry-level set up fee?
No

Add comparison