Apache Pig vs. Hive

Overview
ProductRatingMost Used ByProduct SummaryStarting Price
Apache Pig
Score 8.4 out of 10
N/A
Apache Pig is a programming tool for creating MapReduce programs used in Hadoop.N/A
Hive
Score 8.4 out of 10
N/A
Hive Technology offers their eponymous project management and process management application, providing integrations with many popularly used applications for productivity, cloud storage, and collaboration.
$0
Pricing
Apache PigHive
Editions & Modules
No answers on this topic
Free
$0
Lite
$24
per month per user
Growth
$34
per month per user
Pro
$59
per month per user
Elite
Contact Sales
Offerings
Pricing Offerings
Apache PigHive
Free Trial
NoYes
Free/Freemium Version
YesYes
Premium Consulting/Integration Services
NoNo
Entry-level Setup FeeNo setup feeNo setup fee
Additional DetailsA discount is offered for annual pricing.
More Pricing Information
Community Pulse
Apache PigHive
Considered Both Products
Apache Pig
Chose Apache Pig
Apache Hadoop, Azure Data Lake Storage, Amazon EMR (Elastic MapReduce), Presto (formerly Presto DB), Confluent Platform and Alteryx
Chose Apache Pig
It takes me less time to write a Pig script than get a Spark program running for batch ETL workloads. Compared to Spark, Pig has a steeper learning curve because it employs a proprietary programming language. In one script and one fine, it can handle both Map Reduce and Hadoop. …
Chose Apache Pig
It can accommodate Map Reduce in a single script and a single fine. IT has very much documentation present for easy learning. SQL like queries makes it easy to understand
Chose Apache Pig
Apache Pig might help to start things faster at first and it was one of the best tool years back but it lacks important features that are needed in the data engineering world right now. Pig also has a steeper learning curve since it uses a proprietary language compared to Spark …
Chose Apache Pig
Pig is more focused on scripting in its own PigLatin language rather than integrate into another language like Java/Scala/Python/SQL.
However, for batch ETL workloads, I find that I can write a Pig script quicker than setting up and deploying a Spark program, for example.
Chose Apache Pig
Apache Pig is picked up quickly and can be implemented with very little coding skills. Also the other languages require exact matching of versions during installations which made them somewhat less user-friendly. Also most of the tasks that are done in map reduce can be done …
Chose Apache Pig
I use both Apache Pig and its alternatives like Apache Spark & Apache Hive. Apache Pig was one of the best options in Big Data's initial stages. But now alternatives have taken over the market, rendering Apache Pig behind in the competition. But it is still a better alternative …
Chose Apache Pig
Early on Apache Pig was a great tool for easily writing distributed processing applications without needing to write a complete Java MapReduce job from scratch, but as time as moved on there now better alternatives to get results faster for both ad-hoc analysis and for …
Chose Apache Pig
- Provided better ways for optimized hadoop jobs than Hive but not anymore.
- Spark DSL is much more advanced and compute times are significantly less.
Hive
Chose Hive
In my experience, Hive is better for the work and not so good for financial. That's why we use monday.
Chose Hive
One key difference between Hive and Spark is the way they process data. Hive is a batch-oriented system, which means that it is designed to process large amounts of data in a batch mode rather than in real-time.
In contrast, Spark is a real-time processing platform that is …
Chose Hive
Hive's layout was much smoother and nicer to use.
Chose Hive
More user-friendly. Able to quickly get users adopted and utilizing the platform versus Planview. More intuitive, especially for the user that is not familiar with project management software. This platform was built for everyday users.
Chose Hive
Hive is a bit different than Jira and Monday, which I used mostly. Overall does a great job managing project and helps with team communication. Removes dependency of asking team members for updates by going to conference rooms. With Hive, the team updates the status, and we …
Chose Hive
Hive did what these other tools do. It has Kanban boards, Gantt views, timeline views, reporting, task management, and file uploads. While it is not as feature rich at the lowest subscription level as some of these others, its interface is quite a bit less overwhelming than say …
Chose Hive
  1. Easier to deploy
  2. Better UI/UX
  3. Easy to customize
Chose Hive
Hive for me felt more complex and granular in comparison to other competitors which was a good thing. I enjoyed the layout of viewing projects, the way it integrated timesheets, resourcing, and budgets together, and worked really well to help track episodes and projects.
For …
Chose Hive
So far Hive is the total package for our needs. Offering request forms and proofing/approval out of the box without third party integrations has been a huge upgrade for us along with incredibly reasonable pricing. The support for onboarding has been fantastic and we haven't …
Chose Hive
I would say that in comparison to Asana, Hive is a better interface an UI. I think Asana is more robust in terms of what it can do in conjunction with Confluence but I think Hive is a better entry-level model for new employees. Hive is much simpler and more straight forward and …
Chose Hive
I like Hive better than Trello. Hive is definitely more user-friendly, but Trello had nice shortcuts that I miss in Hive. I would like to organize my board with just one click.
Features
Apache PigHive
Project Management
Comparison of Project Management features of Product A and Product B
Apache Pig
-
Ratings
Hive
7.5
Ratings
2% below category average
Task Management00 Ratings8.30 Ratings
Resource Management00 Ratings7.30 Ratings
Gantt Charts00 Ratings7.70 Ratings
Scheduling00 Ratings7.90 Ratings
Workflow Automation00 Ratings7.50 Ratings
Team Collaboration00 Ratings8.00 Ratings
Support for Agile Methodology00 Ratings8.30 Ratings
Support for Waterfall Methodology00 Ratings7.60 Ratings
Document Management00 Ratings7.10 Ratings
Email integration00 Ratings7.30 Ratings
Mobile Access00 Ratings7.00 Ratings
Timesheet Tracking00 Ratings7.30 Ratings
Change request and Case Management00 Ratings7.00 Ratings
Budget and Expense Management00 Ratings6.60 Ratings
Professional Services Automation
Comparison of Professional Services Automation features of Product A and Product B
Apache Pig
-
Ratings
Hive
7.2
Ratings
5% below category average
Quotes/estimates00 Ratings6.90 Ratings
Invoicing00 Ratings7.20 Ratings
Project & financial reporting00 Ratings7.80 Ratings
Integration with accounting software00 Ratings6.80 Ratings
Best Alternatives
Apache PigHive
Small Businesses

No answers on this topic

Stackby
Stackby
Score 9.0 out of 10
Medium-sized Companies
Cloudera Manager
Cloudera Manager
Score 9.9 out of 10
InEight
InEight
Score 8.3 out of 10
Enterprises
IBM Analytics Engine
IBM Analytics Engine
Score 7.1 out of 10
InEight
InEight
Score 8.3 out of 10
All AlternativesView all alternativesView all alternatives
User Ratings
Apache PigHive
Likelihood to Recommend
8.2
(0 ratings)
8.4
(0 ratings)
Usability
10.0
(0 ratings)
-
(0 ratings)
Support Rating
6.0
(0 ratings)
9.4
(0 ratings)
User Testimonials
Apache PigHive
Likelihood to Recommend
Apache Pig is best suited for ETL-based data processes. It is good in performance in handling and analyzing a large amount of data. it gives faster results than any other similar tool. It is easy to implement and any user with some initial training or some prior SQL knowledge can work on it. Apache Pig is proud to have a large community base globally.
Read full review
Hive is great for managing projects with your team. Assigning tasks is simple enough using Hive. It helps manage team goals for the projects. We are able to create reports (via the dashboard) for the progress and updates to provide to the team based on completed stages. Works great for bigger projects.
Read full review
Pros
  • Iterative Development - you can write aliases/variables, which are not immediately executed and these are stored in a DAG, which is only evaluated upon dumping or storing another alias.
  • Fast execution - Works with MapReduce, Tez, or Spark execution frameworks to provide fast run times at large scales.
  • Local and remote interoperability - Scripts that depend on testing a small dataset locally before moving to the full thing can simply be done with "pig -x local."
Read full review
  • Data warehousing: Hive is often used as a data warehousing platform, allowing users to store and analyze large amounts of structured and semi-structured data. It is especially good at handling data that is too large to be stored and analyzed on a single machine, and supports a wide variety of data formats.
  • Batch processing: Hive is designed for batch processing of large datasets, making it well-suited for tasks such as data ETL (extract, transform, load), data cleansing, and data aggregation.
  • Data transformation: Hive allows users to perform data transformations and manipulations using custom scripts written in Java, Python, or other programming languages. This can be useful for tasks such as data cleansing, data aggregation, and data transformation.
  • Integration with other tools: Hive integrates with a wide variety of other tools and services in the Hadoop ecosystem, such as Pig, Spark, and HBase, allowing users to perform a wide range of data analysis and management tasks.
Read full review
Cons
  • May not fit every need and a SQL-like abstraction may be more effective for some tasks (look at Spark-SQL, Hive, or even an actual DBMS)
  • All Pig jobs are written in a Domain Specific Language so not a lot of transferable knowledge
  • Writing your own User Defined Functions (UDFS) is a nice feature but can be painful to implement in practice
Read full review
  • Organizing tasks by assignees could be better. It's a little cumbersome to check off each person you want. Can you group these?
  • I don't really use any view besides task view. Is there something better I could be using?
  • It would be nice if attachments showed up in a nicer format, maybe with a preview?
Read full review
Usability
It is quick, fast and easy to implement Apache Pig which makes is quite popular to be used.
Read full review
Its a easy tool, the best way to organize the workflow but has room for more improvements.
Read full review
Support Rating
The documentation is adequate. I'm not sure how large of an external community there is for support.
Read full review
Our CSR is easily accessible and they have support built into the app itself. They also have a pretty robust support site. We also took advantage of the free trial and learned so much by putting Hive through the paces and figuring out the best way to mold it to our needs.
Read full review
Alternatives Considered
It takes me less time to write a Pig script than get a Spark program running for batch ETL workloads. Compared to Spark, Pig has a steeper learning curve because it employs a proprietary programming language. In one script and one fine, it can handle both Map Reduce and Hadoop. It has a large amount of documentation available to make learning more convenient.
Read full review
One key difference between Hive and Spark is the way they process data. Hive is a batch-oriented system, which means that it is designed to process large amounts of data in a batch mode rather than in real-time. In contrast, Spark is a real-time processing platform that is designed to handle streaming data and support interactive queries. Another difference is the way they execute queries. Hive uses a SQL-like query language called HiveQL, while Spark supports a wide range of languages and APIs, including SQL, Python, Scala, and R. But we chose Hive due to its simple queries on large datasets and for data warehousing tasks.
Read full review
Return on Investment
  • Return on Investments are significant considering what it can do with traditional analysis techniques. But, other alternatives like Apache Spark, Hive being more efficient, it is hard to stick to Apache Pig.
  • It can handle large datasets pretty easily compared to SQL. But, again, alternatives are more efficient.
  • While working on unstructured, decentralized dataset, Pig is highly beneficial, as it is not a complete deviation from SQL, but it does not take you in complexity MapReduce as well.
Read full review
  • I've gotten to know my colleagues better, knowing their roles makes it faster to contact them to complete tasks and that speed makes us optimize and earn better results
  • The jobs speed made us focus on optimization and customization for the client, and that in a better treatment by the client and better revenue
  • We can understand which tasks takes more time and to stimate better what we can ask for
Read full review
ScreenShots

Hive Screenshots

Screenshot of HIver Technology