Apache Hive is database/data warehouse software that supports data querying and analysis of large datasets stored in the Hadoop distributed file system (HDFS) and other compatible systems, and is distributed under an open source license.
N/A
Oracle Exadata
Score 9.8 out of 10
N/A
Oracle Exadata is an enterprise database platform that runs Oracle Database workloads of any scale and criticality with high performance, availability, and security. Exadata’s scale-out design employs optimizations that let transaction processing, analytics, machine learning, and mixed workloads run faster. Consolidating diverse Oracle Database workloads on Exadata platforms in enterprise data centers, Oracle Cloud Infrastructure (OCI), and multicloud environments helps organizations increase…
$2.90
Per Unit
Presto
Score 10.0 out of 10
N/A
Presto is an open source SQL query engine designed to run queries on data stored in Hadoop or in traditional databases.
Teradata supported development of Presto followed the acquisition of Hadapt and Revelytix.
N/A
Pricing
Apache Hive
Oracle Exadata
Presto
Editions & Modules
No answers on this topic
Database Server
$2.9032
Per Unit
Quarter Rack
$14.5162
Per Unit
No answers on this topic
Offerings
Pricing Offerings
Apache Hive
Oracle Exadata
Presto
Free Trial
No
No
No
Free/Freemium Version
No
No
No
Premium Consulting/Integration Services
No
No
No
Entry-level Setup Fee
No setup fee
No setup fee
No setup fee
Additional Details
—
—
—
More Pricing Information
Community Pulse
Apache Hive
Oracle Exadata
Presto
Considered Multiple Products
Apache Hive
Verified User
Analyst
Chose Apache Hive
Presto is slightly less reliable but much faster for interactive querying. These tools would not be replacements for each other, but rather complements.
We selected Hive because it supports SQL, schema and provides structure on top of hadoop. Having data structured has its benefits, especially if there are thousands of users processing on the same data over and over again. Pig provides the ability to process unstructured data. …
One of the major advantages of using Presto or the main reason why people use Presto (Teradata) is due to that fact it can support multiple data sources - which is lacking as in the case of Apache Hive. But still, most people who come from a Structured data-based background …
Community support and ease of use -not deployment.
It enables querying and analyzing large amounts of data stored in HDFS, on the petabyte scale. It has a query language called HQL that transforms SQL queries into MapReduce jobs that run on Hadoop, and it is wonderful for the …
Hive was one of the first SQL on Hadoop technologies, and it comes bundled with the main Hadoop distributions of HDP and CDH. Since its release, it has gained good improvements, but selecting the right SQL on Hadoop technology requires a good understanding of the strengths and …
I think Presto is one of the best solutions out there today at the cutting edge for interactive query analysis. One of the challenges is presto is a niche tool for the interactive query use case and doesn't have the knobs and whistles as much as Spark. In the foreseeable future …
Software work execution is on a large scale, it is good to use for new projects or organizational changes, data lineage mapping has always been dubious but this one has had good results. You can store and synchronize data from different departments, the storage process can be manual but it is best automated.
Oracle Exadata is well-suited for environments where massive performance for Oracle databases is required. Storage indexes reduce the unnecessary I/O. Smart Flash Cache accelerates random reads/writes.
Our OLTP application demands very high concurrency. Multi-node Exadata provides high availability and zero downtime during DB patching. It comes with lots of built-in automations, so it reduces many routine tasks for sysadmins, like network, storage, and VM configuration, and it also reduces many Oracle DBA tasks, like Oracle software installation, patching, and upgrades.
Presto is for interactive simple queries, where Hive is for reliable processing. If you have a fact-dim join, presto is great..however for fact-fact joins presto is not the solution.. Presto is a great replacement for proprietary technology like Vertica
Apache Hive allows use to write expressive solutions to complex problems thanks to its SQL-like syntax.
Relatively easy to set up and start using.
Very little ramp-up to start using the actual product, documentation is very thorough, there is an active community, and the code base is constantly being improved.
Oracle Database : Deliver industry-leading security, high availability and scalability with Oracle Database, which has been significantly enhanced to take advantage of the Oracle Exadata Storage Servers.
Exadata Smart Scan : Improve query performance by offloading intensive query processing and data mining scoring to scalable intelligent storage servers.
Smart Flash Cache : Transparently cache 'hot' read and write data to fast solid-state storage, improving query response times and throughput. Exadata systems use the latest PCI flash technology rather than flash disks. PCI flash delivers ultra-high performance by placing flash directly on the high speed PCI bus rather than behind slow disk controllers.
Hybrid Columnar Compression : Reduce the size of data warehousing tables by 10x, and archive tables by 50x, to improve performance and lower storage costs for primary, standby, and backup databases. Query high, query low, archive high and archive low.
Infiniband Network : Connect multiple Oracle Exadata Database Machines using the InfiniBand fabric to form a larger single system image configuration. Each InfiniBand link provides 40 Gigabits of bandwidth–many times higher than traditional storage or server networks.
Petabyte Scalability : Easily scale data warehouse to support enterprise data growth.
Linking, embedding links and adding images is easy enough.
Once you have become familiar with the interface, Presto becomes very quick & easy to use (but, you have to practice & repeat to know what you are doing - it is not as intuitive as one would hope).
Organizing & design is fairly simple with click & drag parameters.
The process of patching and upgrade of Exadata server components could be improved with a goal to minimize the overall effort, make it fully automated and transparent.
Improved guidelines and possibly more sophisticated tools for sizing of new Exadata servers for migration from old legacy hardware.
Presto was not designed for large fact fact joins. This is by design as presto does not leverage disk and used memory for processing which in turn makes it fast.. However, this is a tradeoff..in an ideal world, people would like to use one system for all their use cases, and presto should get exhaustive by solving this problem.
Resource allocation is not similar to YARN and presto has a priority queue based query resource allocation..so a query that takes long takes longer...this might be alleviated by giving some more control back to the user to define priority/override.
UDF Support is not available in presto. You will have to write your own functions..while this is good for performance, it comes at a huge overhead of building exclusively for presto and not being interoperable with other systems like Hive, SparkSQL etc.
Hive is a very good big data analysis and ad-hoc query platform, which supports scaling also. The BI processes can be easily integrated with Hadoop via the Hive. It can deal with a much larger data set that traditional RDBMS can not. It is a "must-have" component of the big data domain.
I am comparing Exadata with the Oracle RAC database experience. In addition to Oracle RAC features, Exadata provides automatic performance optimization through Smart Scan and storage indexes. Deep integration with the Oracle ecosystem and tight coupling with Oracle Enterprise Manager for monitoring and management. Some downsides of Exadata are: a steep learning curve, concepts like cell offloading, IORM, and flash cache behavior aren’t intuitive initially. Operating Exadata requires specialized DBA skills.
Apache Hive is a FOSS project and its open source. We need not definitely comment on anything about the support of open source and its developer community. But, it has got tremendous developer support, awesome documentation. I would justify the fact that much support can be gathered from the community backup.
Besides Hive, I have used Google BigQuery, which is costly but have very high computation speed. Amazon Redshift is the another product, I used in my recent organisation. Both Redshift and BigQuery are managed solution whereas Hive needs to be managed
Oracle Exadata Database Machine had the best performance overall hands down. It clearly beat the competition and we were seeing 1000X improvement on SAP HANA. Oracle Exadata Database Machine beat that without us refactoring our code. To achieve that in HANA, we had to refactor the code somewhat. Now this was for our limited POC of 5 use cases. Given the large number of stored procedures we had in Sybase, we need to capture more production metrics but we are seeing incredible performance.
Presto is good for a templated design appeal. You cannot be too creative via this interface - but, the layout and options make the finalized visual product appealing to customers. The other design products I use are for different purposes and not really comparable to Presto.
Single support from a single vendor with both machine and database from Oracle, which is costing us less.
With Exadata, we need less technical manpower and less technical support. A business transaction with the integrated and centralized database helps us focus on other business needs.
We don't need to buy additional licenses and Hardware for the next 3 to 5 years.