Apache Hive is database/data warehouse software that supports data querying and analysis of large datasets stored in the Hadoop distributed file system (HDFS) and other compatible systems, and is distributed under an open source license.
N/A
Oracle Exadata
Score 10.0 out of 10
N/A
Oracle Exadata is an enterprise database platform that runs Oracle Database workloads of any scale and criticality with high performance, availability, and security. Exadata’s scale-out design employs optimizations that let transaction processing, analytics, machine learning, and mixed workloads run faster. Consolidating diverse Oracle Database workloads on Exadata platforms in enterprise data centers, Oracle Cloud Infrastructure (OCI), and multicloud environments helps organizations increase…
$2.90
Per Unit
Pricing
Apache Hive
Oracle Exadata
Editions & Modules
No answers on this topic
Database Server
$2.9032
Per Unit
Quarter Rack
$14.5162
Per Unit
Offerings
Pricing Offerings
Apache Hive
Oracle Exadata
Free Trial
No
No
Free/Freemium Version
No
No
Premium Consulting/Integration Services
No
No
Entry-level Setup Fee
No setup fee
No setup fee
Additional Details
—
—
More Pricing Information
Community Pulse
Apache Hive
Oracle Exadata
Considered Both Products
Apache Hive
Verified User
Anonymous
Chose Apache Hive
To query a huge, distributed dataset, Apache Hive was built by Facebook. Unlike Apache Hive, Apache Spark is an in-memory computation engine, which is why it is significantly quicker than Apache Hive at querying large amounts of data. In contrast to Apache HBase, Apache Hive is …
Community support and ease of use -not deployment.
It enables querying and analyzing large amounts of data stored in HDFS, on the petabyte scale. It has a query language called HQL that transforms SQL queries into MapReduce jobs that run on Hadoop, and it is wonderful for the …
Apache Spark is similar in the sense that it too can be used to query and process large amounts of data through its Dataframe interface. Hive is better for short-term querying while Spark is better for persistent and long-term analysis. Another product is Impala. For our …
We have used a simple but necessary function such as merging certain data tables, which although they may be from different areas, complement each other or are necessary, you can use metadata if what you need is to validate the origin of your information and what impact it has, …
Apache Hadoop is built on top of the Hadoop File system so it gives its best when integrated with Hadoop. Data analysis and query optimization become very easy when used with Hadoop to perform Extract transform load operations. As Hadoop is a big data system and handles large …
We have used the system to migrate data either for new versions or because we will use another operating program, the software helps us to synchronize programs between different operating systems, a history of information can be kept constant, it can be sent to third parties …
Queries are easy to write and interface is similar to SQL so learning overhead is reduced. Multi user and data type support is provided. Can be easily scaled for very large amount of analytics. It is very flexible in terms of using file formats.
Apache Hive is a query language developed by Facebook to query over a large distributed dataset. Apache is a query engine that runs on top of HDFS, so it utilizes the resources of HDFS Hadoop setup, while Apache Spark is an in memory compute engine, and that's why [it is] much …
Besides Hive, I have used Google BigQuery, which is costly but have very high computation speed. Amazon Redshift is the another product, I used in my recent organisation. Both Redshift and BigQuery are managed solution whereas Hive needs to be managed
Hive and Spark have the same parent company hence they share a lot of common features. Hive follows SQL syntax while Spark has support for RDD, DataFrame API. DataFrame API supports both SQL syntax and has custom functions to perform the same functionality. Spark is faster and …
One of the major advantages of using Presto or the main reason why people use Presto (Teradata) is due to that fact it can support multiple data sources - which is lacking as in the case of Apache Hive. But still, most people who come from a Structured data-based background …
Easy to understand, well supported by the community, good documentation. However, it is possible that SAP Business Warehouse could be a good fit, too, even maybe better. I did not have the chance to try it though. We selected Apache Hive because it was far less expensive and …
For storing bulk amount of data in a tabular manner, and where there's no need need of primary key, or just in case, if redundant data is received, it will not cause a problem. For small amounts of data, it does run MR, so beware. If your intention is to use it as a …
I wasn't part of the evaluation process for Apache Hive. This was already implemented when I joined the company. I have worked with other big data plaftforms and I personally thinks most of them are quite comporable to one another. It really depends on what the company is going …
Apache Pig is probably the most direct technology to compare to Hive and has several different use cases to Hive. If you want to simplify processing tasks that run using MapReduce then Apache Pig may be a better tool for the job. However if you are going to be running many …
A unique architecture of Oracle Exadata machine which consists of several components: compute, storage cells with offloaded SQL processing within the cell, smart cache. In addition it is an Oracle RAC server with high speed interconnect between its built-in nodes.
Oracle Database Exadata Cloud Service allocates built-in cloud automation and enhances enterprise-class business continuity by enhancing zero downtime maintenance which is contrary to other alternatives such as Apache Hive.
Vice President, Chief Architect, Development Manager and Software Engineer
Chose Oracle Exadata
Oracle Exadata Database Machine had the best performance overall hands down. It clearly beat the competition and we were seeing 1000X improvement on SAP HANA. Oracle Exadata Database Machine beat that without us refactoring our code. To achieve that in HANA, we had to …
IBM POWER System is a general purpose hardware optimize to runs various software with high-performance resource intensive operations. On the other hand, Oracle Exadata Database Machine is specifically engineered to run Oracle Database software efficiently, this combination of …
IBM AIX and HP-UX implementations of Oracle database solutions have a lot of performance issues. Both do not provide as much robust configuration customization as Exadata. Hardware support is limited. There is generally a long delay between hardware update being certified with …
We have done a proof of concept for both and have seen a lift with our batch processing and all other aspects with Oracle Exadata. Oracle Exadata storage servers have been playing a key role with the overall success compared to other products.
We selected it just from a performance perspective, and that the ROI with the Oracle Exadata Database Machine is bigger than other machines. You can run with it for at least 5 years.
I do not think there is any alternative for Exadata. Flash storage or SSD can not solve the IO bottleneck issues the way Exadata handles the IO subsystem.
Exadata beats the competition because the smart scan and offloading technology is more about software than hardware, so you cannot just buy a beefy server and add flash disks to compete. The Exadata software is what makes it special.
Apache Hive shines for ad-hoc analysis and plugging into BI tools. Its SQL-like syntax allows for ease of use not for only for engineers but also for data analysts. Through our experience, there are probably more desirable tools to use if you are planning on integrating Hive into your processing pipeline.
First, get the database on Oracle. If you are in an Oracle stack, it would be much better to use the Oracle products. If you are driving a Ferrari, you wouldn’t put a Mercedes engine in it. If you are writing a query, you cannot rely on other brands. Since I'm an architect, when I look for a product, I look for performance.
The installation is easy because it comes out-of-the-box and you just start using it.
Previous to Oracle Exadata, we were using a normal Oracle RAC service. We were just waiting for this product to come out.
I'm currently writing a data warehouse on Exadata. Before this solution, we were aiming for this to be completed by 8 a.m., when our ETLs would finish. With the help of Exadata's special features, this was reduced to 3 a.m. This solution allows us to bring more data within the same time period. It provides us with more subject areas that provide more reports to our users. Our ETL times reduced to 65%, then to 50%.
Customize-able for specific functionality optimized for combination of online transaction or analytical processing.
Ability to serve mix workloads with resource management feature enables prioritizing allocation for certain workload.
Scale-able on-premise with compatibility for cloud deployment offers flexible solution for organization considering to transition from on-premise solution.
Hive is a very good big data analysis and ad-hoc query platform, which supports scaling also. The BI processes can be easily integrated with Hadoop via the Hive. It can deal with a much larger data set that traditional RDBMS can not. It is a "must-have" component of the big data domain.
Apache Hive is a FOSS project and its open source. We need not definitely comment on anything about the support of open source and its developer community. But, it has got tremendous developer support, awesome documentation. I would justify the fact that much support can be gathered from the community backup.
We have used a simple but necessary function such as merging certain data tables, which although they may be from different areas, complement each other or are necessary, you can use metadata if what you need is to validate the origin of your information and what impact it has, is also feasible.
A unique architecture of Oracle Exadata machine which consists of several components: compute, storage cells with offloaded SQL processing within the cell, smart cache. In addition it is an Oracle RAC server with high speed interconnect between its built-in nodes.
Single support from a single vendor with both machine and database from Oracle, which is costing us less.
With Exadata, we need less technical manpower and less technical support. A business transaction with the integrated and centralized database helps us focus on other business needs.
We don't need to buy additional licenses and Hardware for the next 3 to 5 years.