Apache Hive is database/data warehouse software that supports data querying and analysis of large datasets stored in the Hadoop distributed file system (HDFS) and other compatible systems, and is distributed under an open source license.
N/A
Snowflake
Score 8.9 out of 10
N/A
The Snowflake Cloud Data Platform is the eponymous data warehouse with, from the company in San Mateo, a cloud and SQL based DW that aims to allow users to unify, integrate, analyze, and share previously siloed data in secure, governed, and compliant ways. With it, users can securely access the Data Cloud to share live data with customers and business partners, and connect with other organizations doing business as data consumers, data providers, and data service providers.
N/A
Pricing
Apache Hive
Snowflake
Editions & Modules
No answers on this topic
No answers on this topic
Offerings
Pricing Offerings
Apache Hive
Snowflake
Free Trial
No
Yes
Free/Freemium Version
No
No
Premium Consulting/Integration Services
No
No
Entry-level Setup Fee
No setup fee
No setup fee
Additional Details
—
—
More Pricing Information
Community Pulse
Apache Hive
Snowflake
Considered Both Products
Apache Hive
Verified User
Anonymous
Chose Apache Hive
To query a huge, distributed dataset, Apache Hive was built by Facebook. Unlike Apache Hive, Apache Spark is an in-memory computation engine, which is why it is significantly quicker than Apache Hive at querying large amounts of data. In contrast to Apache HBase, Apache Hive is …
Community support and ease of use -not deployment.
It enables querying and analyzing large amounts of data stored in HDFS, on the petabyte scale. It has a query language called HQL that transforms SQL queries into MapReduce jobs that run on Hadoop, and it is wonderful for the …
Apache Spark is similar in the sense that it too can be used to query and process large amounts of data through its Dataframe interface. Hive is better for short-term querying while Spark is better for persistent and long-term analysis. Another product is Impala. For our …
We have used a simple but necessary function such as merging certain data tables, which although they may be from different areas, complement each other or are necessary, you can use metadata if what you need is to validate the origin of your information and what impact it has, …
Apache Hadoop is built on top of the Hadoop File system so it gives its best when integrated with Hadoop. Data analysis and query optimization become very easy when used with Hadoop to perform Extract transform load operations. As Hadoop is a big data system and handles large …
We have used the system to migrate data either for new versions or because we will use another operating program, the software helps us to synchronize programs between different operating systems, a history of information can be kept constant, it can be sent to third parties …
Queries are easy to write and interface is similar to SQL so learning overhead is reduced. Multi user and data type support is provided. Can be easily scaled for very large amount of analytics. It is very flexible in terms of using file formats.
Apache Hive is a query language developed by Facebook to query over a large distributed dataset. Apache is a query engine that runs on top of HDFS, so it utilizes the resources of HDFS Hadoop setup, while Apache Spark is an in memory compute engine, and that's why [it is] much …
Besides Hive, I have used Google BigQuery, which is costly but have very high computation speed. Amazon Redshift is the another product, I used in my recent organisation. Both Redshift and BigQuery are managed solution whereas Hive needs to be managed
Hive and Spark have the same parent company hence they share a lot of common features. Hive follows SQL syntax while Spark has support for RDD, DataFrame API. DataFrame API supports both SQL syntax and has custom functions to perform the same functionality. Spark is faster and …
One of the major advantages of using Presto or the main reason why people use Presto (Teradata) is due to that fact it can support multiple data sources - which is lacking as in the case of Apache Hive. But still, most people who come from a Structured data-based background …
Easy to understand, well supported by the community, good documentation. However, it is possible that SAP Business Warehouse could be a good fit, too, even maybe better. I did not have the chance to try it though. We selected Apache Hive because it was far less expensive and …
For storing bulk amount of data in a tabular manner, and where there's no need need of primary key, or just in case, if redundant data is received, it will not cause a problem. For small amounts of data, it does run MR, so beware. If your intention is to use it as a …
I wasn't part of the evaluation process for Apache Hive. This was already implemented when I joined the company. I have worked with other big data plaftforms and I personally thinks most of them are quite comporable to one another. It really depends on what the company is going …
Apache Pig is probably the most direct technology to compare to Hive and has several different use cases to Hive. If you want to simplify processing tasks that run using MapReduce then Apache Pig may be a better tool for the job. However if you are going to be running many …
Snowflake provides various features, such as integration with Python using Snowpark. The reporting feature that caters to your small reporting needs is Snowsight. The Snowflake data marketplace is where you can get multiple data for free and even some of the data which you can …
These are comparable products that can make sense depending on the specific needs of your organization. All are certainly serviceable and have varying pros and cons. Snowflake seems to provide the greatest degree of flexibility and easy scalability as new data gets brought into …
We needed scalability and a new way of organizing our data; Snowflake allowed us to have a clearer view of our data warehouses and schemas. Snowflake is also way superior in terms of speed and quick insights from the raw data you query, which is very valuable to us.
Snowflake has an attractive pricing model with auto-suspend and auto-resume and pay per use. AWS Redshift requires higher administrative efforts to maintain and scale the platform whereas with Snowflake those admin tasks are not needed or automatically taken care of.
We had a MS SQL server with over 2 TB of ram & 51 processors that we were using, that could no longer handle our workload. Snowflake can handle 3 times that workload with ease and efficiency.
Snowflake is much faster and easier to write queries and pull data. But the visualization part of Snowflake is not as good as them. Also, Snowflake only supports SQL queries but not python or other languages. So basically Snowflake is the expert in its field but not suitable …
We particularly liked Snowflake's security model as well as its unique storage (whereby everything is essentially a pointer to immutable micro-partitions, which is the key behind its zero-copy cloning, its secure sharing, its time travel, etc.). and also how it separates …
While Snowflake is more open to cloud eco system, SAP integrated well with SAP eco system products like SAP ECC or SAP S/4. So for people who have invested heavily in SAP eco system including SAP ECC or S/4, it makes sense to go with SAP DWC which is also evolving very rapidly. …
In my opinion, the other tools have similar and some different features; however, when I ran proof of technologies between Synapse and Snowflake. Snowflake did things better or just had functionality that the other tools did not. One that stuck out at the time was scale up …
Each of the other solutions were cloud vendor specific, Snowflake can ride on either Amazon Web Services, Microsoft Azure, or Google Cloud. The fact that they are ANSI-sql compliant and have an effective means of offloading data makes them portable and easy to sell to teams …
Azure and Snowflake compared very similarly, but Snowflake provided more options to integrate and connect with tools/companies that were not partners. It seemed to be a more flexible environment. The barrier for entry on Oracle and Google we just too complicated. In particular, …
I have had the experience of using one more database management system at my previous workplace. What Snowflake provides is better user-friendly consoles, suggestions while writing a query, ease of access to connect to various BI platforms to analyze, [and a] more robust system …
Snowflake has won the match because it is giving an excellent performance with its efficient features and reliable results. This is a totally secure program for our precious and important data.
Our initial data warehousing solution was Treasure Data. We had issues with the costly pricing model, which would be exhorbitant if we want to hold our data in memory and query using Presto. As a result, some heavy lifting was done in Hive (managed by Treasure Data); …
In my experience running the data management practice at InterWorks, we believe that cloud data warehouse products will eventually serve the majority of data warehousing use cases and power data analytics at most companies. Of this cohort, we believe that Snowflake is the best …
Redshift compute and storage can be scaled up/down together (though they added some features recently, they don't quite add up). I haven't tried Avalanche or Firebolt but would love to in the near future, due to their pedigree or revolutionary billing methods.
- Cost was the main aspect on the decision. - Performance was in par or better compared to other tools in the market. - Snowflake in my opinion stacks better than other tools I have used in the past.
Accommodates future data types such as JSON and XML. Scalability is another advantage. Pay per use is beneficial for organizations like yours. Direct connectors with AWS help us to go with it. No limit on user creation and clone data not eating up extra disk space are a few …
Since we switch from amazon redshift to Snowflake, we found Snowflake is much better than redshift in many ways, including the data integrate and data pull. However, comparing directly pull data from amazon s3, Snowflake is quite slow in terms of data pull speed and the more …
Compared to Amazon Redshift, Snowflake is slightly easier and faster to achieve ROI but based on the user's perspective, the two tools have very little difference since both are leveraging SQL to pull data from AWS S3. Snowflake is also working with Microsoft Azure but it is …
Our issue with Redshift was that it was very expensive. On top of that, queries were still slow and if we used more of Redshift's memory, then it would have cost even more. Snowflake is not cheap, but less costly for us. Plus, the performance was much better. Also, we got to …
Apache Hive shines for ad-hoc analysis and plugging into BI tools. Its SQL-like syntax allows for ease of use not for only for engineers but also for data analysts. Through our experience, there are probably more desirable tools to use if you are planning on integrating Hive into your processing pipeline.
If you need a quick query, snowflake is the way to go. It's super simple and scalable; we were struggling before with Azure, and with Snowflake, everything runs smoothly, and we have more control over our schemas and warehouses. Snowflake, in my opinion, is the next step when you want to scale your business and manage data. If your company is still small, there may be cheaper options.
Snowflake scales appropriately allowing you to manage expense for peak and off peak times for pulling and data retrieval and data centric processing jobs
Snowflake offers a marketplace solution that allows you to sell and subscribe to different data sources
Snowflake manages concurrency better in our trials than other premium competitors
Snowflake has little to no setup and ramp up time
Snowflake offers online training for various employee types
Do not force customers to renew for same or higher amount to avoid loosing unused credits. Already paid credits should not expire (at least within a reasonable time frame), independent of renewal deal size.
SnowFlake is very cost effective and we also like the fact we can stop, start and spin up additional processing engines as we need to. We also like the fact that it's easy to connect our SQL IDEs to Snowflake and write our queries in the environment that we are used to
Hive is a very good big data analysis and ad-hoc query platform, which supports scaling also. The BI processes can be easily integrated with Hadoop via the Hive. It can deal with a much larger data set that traditional RDBMS can not. It is a "must-have" component of the big data domain.
The interface is similar to other SQL query systems I've used and is fairly easy to use. My only complaint is the syntax issues. Another thing is that the error messages are not always the easiest thing to understand, especially when you incorporate temp tables. Some of that is to be expected with any new database.
Apache Hive is a FOSS project and its open source. We need not definitely comment on anything about the support of open source and its developer community. But, it has got tremendous developer support, awesome documentation. I would justify the fact that much support can be gathered from the community backup.
We have had terrific experiences with Snowflake support. They have drilled into queries and given us tremendous detail and helpful answers. In one case they even figured out how a particular product was interacting with Snowflake, via its queries, and gave us detail to go back to that product's vendor because the Snowflake support team identified a fault in its operation. We got it solved without lots of back-and-forth or finger-pointing because the Snowflake team gave such detailed information.
We have used a simple but necessary function such as merging certain data tables, which although they may be from different areas, complement each other or are necessary, you can use metadata if what you need is to validate the origin of your information and what impact it has, is also feasible.
Snowflake provides various features, such as integration with Python using Snowpark. The reporting feature that caters to your small reporting needs is Snowsight. The Snowflake data marketplace is where you can get multiple data for free and even some of the data which you can buy according to your needs. And the integration options with various tools like Sigma are add-ons.