Apache Hadoop vs. Teradata Vantage

Overview
ProductRatingMost Used ByProduct SummaryStarting Price
Hadoop
Score 7.9 out of 10
N/A
Hadoop is an open source software from Apache, supporting distributed processing and data storage. Hadoop is popular for its scalability, reliability, and functionality available across commoditized hardware.N/A
Teradata Vantage
Score 8.2 out of 10
N/A
Teradata Vantage is presented as a modern analytics cloud platform that unifies everything—data lakes, data warehouses, analytics, and new data sources and types. Supports hybrid multi-cloud environments and priced for flexibility, Vantage delivers unlimited intelligence to build the future of business. Users can deploy Vantage on public clouds (such as AWS, Azure, and GCP), hybrid multi-cloud environments, on-premises with Teradata IntelliFlex, or on commodity hardware with VMware.
$4,800
per month
Pricing
Apache HadoopTeradata Vantage
Editions & Modules
No answers on this topic
Teradata VantageCloud Lake
from $4800
per month
Teradata VantageCloud Enterprise
from $9000
per month
Offerings
Pricing Offerings
HadoopTeradata Vantage
Free Trial
NoYes
Free/Freemium Version
YesNo
Premium Consulting/Integration Services
NoNo
Entry-level Setup FeeNo setup feeOptional
Additional Details
More Pricing Information
Community Pulse
Apache HadoopTeradata Vantage
Considered Both Products
Hadoop
Chose Hadoop
It’s open source nature
it’s community support
its being configurable
Chose Hadoop
Different departments of my organization have been getting the benefit from Apache Hadoop as it serves the purpose of saving lives when large amounts of data is unable to be converted and processed in a timely manner from a node or a simple computer. Hadoop also has an easier …
Chose Hadoop
I feel that this is a highly reliable and scalable solution computing technology that is highly capable of processing large data sets across multiple servers and thousands of machines in a well-defined and distributed manner. Apache Hadoop can automatically scale up the number …
Chose Hadoop
Spark is a good alternative to Hadoop that can have faster querying and processing performance and can offer more flexibility in terms of applications that it can support.

Google Bigquery has also been a great alternative and is especially great in terms of ease of use. The …
Chose Hadoop
MariaDB - Better to be already in the cloud you will use it for. Issues have improved as it has matured over the year.s
CockroachDB - Not nearly as performant (even out of the box) as Apache Hadoop. More configurations required just to make it work. In memory cacheing is an issue.
Chose Hadoop
Hadoop utilizes a SQL structure, which is great. You pay less for the services, but it's definitely less of an enterprise-level option and more just a good place to store your seldom-used data. Teradata and AWS are a lot faster in returning queries than Hadoop, but you pay …
Chose Hadoop
Hands down, Hadoop is less expensive than the other platforms we considered. Cloudera was easier to set up but the expense ruled it out. MS-SQL didn't have the performance we saw with the Hadoop clusters and was more expensive. We considered MS-SQL mainly for its ability …
Chose Hadoop
When comparing to the sophistication of IBM GPFS (Spectrum Scale) to Hadoop, it is clear that Spectrum Scale is a much better choice. That is maybe something you don't want to hear, but in all of our research, this has been the final decision of the client.
Chose Hadoop
Apache Spark can be considered as an alternative because of its similar capabilities around processing and storing big data. The reason we went with Hadoop was the literature available online and integration capability with platforms like R Studio. The popularity of Hadoop has …
Chose Hadoop
  • For real-time streaming, use Spark; can provide a stark contrast to the way MR works
  • Hadoop offers a scalable, cost-effective and highly available solution for big data storage and processing.
  • Amazon Redshift is somewhat closer to Hadoop. But to analyze Petabytes of data Hadoop …
Chose Hadoop
Hadoop offers a scalable, cost-effective and highly available solution for big data storage and processing. The use of a non-proprietary physical layer greatly reduces dependency on technology. It also offers elastic dimensioning capability when deployed on virtual machines or …
Chose Hadoop
I haven't worked with other Big Data aggregation services like Hadoop. As far as I know, Hadoop is the leading choice in this field with good cause. There is a lot of community support, custom modules, paid consultants, free and paid training. All this makes it an ideal choice …
Chose Hadoop
No SQL database were evaluated along with MPP platform. Hadoop performs very well compared to the other platforms. Also since lot of investment goes into Hadoop there is a good chance of getting what one needs from the developer community.
Chose Hadoop
Amazon Redshift is some what closer to Hadoop. But to analyze Petabytes of data Hadoop as better performance.
Chose Hadoop
As I am new to the hadoop ecosystem I have not used or evaluated any other similar products at this time. This was handed to me from a previous much older installation that was very under utilized. Our new platform will be working the new cluster much harder with jobs that run …
Chose Hadoop
Hadoop was a cheaper alternative to Amazon. Since I had to pay for every minute I use with Amazon, I had to make sure multiple times that the code was good enough before I purchased with Amazon. But since Hadoop was available on the cluster, I had the opportunity to code on the …
Chose Hadoop
Hadoop being open source, is cheaper to use and do POCs for clients. Cloudera, Hortonworks and MapR also compete to contribute to open source Hadoop and keep their product conceptually similar to Hadoop.
Chose Hadoop
Apache Spark has an in memory processing model, making it powerful for lightning fast data processing. Apache Spark also exposes Scala and Python in APIs which is one of the most commonly used programming languages in data analytic and data processing domains.
Chose Hadoop
Not used any other product than Hadoop and I don't think our company will switch to any other product, as Hadoop is providing excellent results. Our company is growing rapidly, Hadoop helps to keep up our performance and meet customer expectations. We also use HDFS which …
Chose Hadoop
Hadoop provides storage for large data sets and a powerful processing model to crunch and transform huge amounts of data. It does not assume the underlying hardware or infrastructure and enables the users to build data processing infrastructure from commodity hardware. All the …
Chose Hadoop
Processing of big data has been the ultimate need for the me choosing Hadoop. Big data is massive and messy, and it’s coming at you uncontrolled. Data are gathered to be analyzed to discover patterns and correlations that could not be initially apparent, but might be useful in …
Chose Hadoop
Hadoop solves lot of problems (involving unstructured data and huge volumes of data ) better than traditional database systems . And it is completely free and open source ( so lots of cost savings ). Data analysis is very fast when compared to old systems, resulting in more …
Teradata Vantage
Chose Teradata Vantage
I had no experience with other similar products to be able to compare

No tuve experiencia sobre otros productos similares como para poder comparar
Chose Teradata Vantage
Son similares. No seleccioné Teradata Vantage. Cuando ingrese a la compañía ya existe Teradata Vantage como herramienta corporativa They are similar. I did not choose Teradata Vantage. When I joined the company, Teradata Vantage was already established as a corporate tool.
Chose Teradata Vantage
Performance and capacilities in order to manage high volumes of data, multiples joins and complex queries
Chose Teradata Vantage
Uso pre-existente, vinculo con el proveedor. Conocimiento de la solucion. Existing use, link with the provider. Knowledge of the solution.
Chose Teradata Vantage
The Teradata is leader and reference in the market.We had a project to migrate from Teradata on premise to Teradata Cloud, bring advantages por example: we can inprovement our worklouds with low impacts for our infra solution e bring better experience to work in the cloud tools …
Chose Teradata Vantage
The Teradata is leader and reference in the market.
We had a project to migrate from Teradata on premise to Teradata Cloud, bring advantages por example: we can inprovement our worklouds with low impacts for our infra solution e bring better experience to work in the cloud tools …
Chose Teradata Vantage
At this time because it is a system that we are already using and it responds to all our needs.
Chose Teradata Vantage
Teradata Vantage complements the above products, but Teradata Vantage is excellent for data analytics.
Chose Teradata Vantage
Snowflake, Google BigQuery and Oracle Exadata
Chose Teradata Vantage
We already had Teradata onPrem. The main driver for Teradata Vantage has been the simlpicity to move on Cloud
Chose Teradata Vantage
Snowflake felt very limited
Chose Teradata Vantage
Performance of Teradata was fasr superior
Chose Teradata Vantage
I have sought feedback from other people and nothing would appear to do what Vantage does for us. There may be new capabilities that are comparable to other providers, but for the everyday data engineering, I would struggle to find anything better
Chose Teradata Vantage
still not fully adopted, not able to compare
Chose Teradata Vantage
The best all-around option for flexibility, scalability, performance, and manageability.
Chose Teradata Vantage
Teradata is way ahead of its competitor because of its unique features of ensuring data privacy and data never gets corrupted even in worst case scenario. In most cases, the data corruption is a major issue if left unused and it leads to important data being wiped off which in …
Chose Teradata Vantage
Oracle Exadata is an excellent product. Performs mass data processing with similar capability compared to Teradata. Some features Exadata has lack for Teradata Vantage, such as archive generation, consistent reading and writing (simultaneously), RMAN backing up online …
Chose Teradata Vantage
Teradata Vantage and the others are used for diferent kind of uses case, they work together for giving a full service to users and the company.
Chose Teradata Vantage
I have used Databricks and DataLake, which are better with semi-structured data than Teradata; and also integrate with other tools. And the most important factor is that in these tools you can separate storage from processing.
Chose Teradata Vantage
Apache Spark, MySQL, Oracle Database Exadata Cloud Service and Apache Hadoop
Chose Teradata Vantage
To be fair, I didn’t select Teradata. I do think that they are comparable. There are some things SQL server does better than Teradata and vice versa. For example, sql server will underline potential errors while you are coding. It also will auto-populate table and column names …
Chose Teradata Vantage
Teradata is one of the best databases compared to all the other RDBMS databases. It's because of the well thought architecture and features built around it. Nothing can match to its performance and speed. If cost is not the factor, I would highly recommend Teradata to anyone …
Chose Teradata Vantage
There are many alternatives available in the market and many of them are cheaper as well. But you need to be very clear in your mind why you want to go for Teradata, what is the future plan for it and how are you going to make the most of it because it is certainly much more …
Chose Teradata Vantage
I have used almost every metadata product out there. Teradata about the middle of the road as far as I am concerned. It's not great, it's not horrible. Again, as in my previous comments, if you have a full Teradata environment, then go for the Teradata Master Data Management. …
Best Alternatives
Apache HadoopTeradata Vantage
Small Businesses

No answers on this topic

Google BigQuery
Google BigQuery
Score 8.5 out of 10
Medium-sized Companies
Cloudera Manager
Cloudera Manager
Score 9.9 out of 10
Snowflake
Snowflake
Score 8.9 out of 10
Enterprises
IBM Analytics Engine
IBM Analytics Engine
Score 7.1 out of 10
Snowflake
Snowflake
Score 8.9 out of 10
All AlternativesView all alternativesView all alternatives
User Ratings
Apache HadoopTeradata Vantage
Likelihood to Recommend
8.0
(0 ratings)
9.5
(0 ratings)
Likelihood to Renew
9.6
(0 ratings)
7.6
(0 ratings)
Usability
8.0
(0 ratings)
7.8
(0 ratings)
Performance
8.0
(0 ratings)
-
(0 ratings)
Support Rating
7.5
(0 ratings)
8.0
(0 ratings)
Online Training
6.1
(0 ratings)
-
(0 ratings)
User Testimonials
Apache HadoopTeradata Vantage
Likelihood to Recommend
Apache Hadoop (and its subsequent add-ons) are well-suited to larger, unstructured data flows, such as aggregation of web traffic or advertising. Geospatial algorithms and their outputs are well-suited for this kind of aggregation as structuring that data is challenging, but leaving it unstructured and performing queries as-needed is a better fit for most business models. With the advent of data science, I would expect Hadoop fits a LOT of their initial outputs quite well.
Read full review
Teradata Vantage is well suited for large scale ETL pipelines like the ones we developed for anti money laundering risk matrices. It handles heavy joins, aggregations, and transformations on transactional data efficiently. We generate alert variables, adjust for inflation, and monitor establishments monthly with it, all integrated with Python and Control-M for a centralised automation across the company. For less appropriate, I would say that heavy resource demands might slow down experimentation for iterative work.
Read full review
Pros
  • HDFS is reliable and solid, and in my experience with it, there are very few problems using it
  • Enterprise support from different vendors makes it easier to 'sell' inside an enterprise
  • It provides High Scalability and Redundancy
  • Horizontal scaling and distributed architecture
Read full review
  • ETL (Extract - Transfor - Load)
  • NOS to send data from Teradata Vantage to S3 and from S3 to Teradata Vantage
  • Teradata GeoSpacial feature
  • Bulk reading and writing in huge tables
  • MPP capacity already mature
  • Temporal Capacity more mature that other solutions
  • TASM
Read full review
Cons
  • Hadoop is a batch oriented processing framework, it lacks real time or stream processing.
  • Hadoop's HDFS file system is not a POSIX compliant file system and does not work well with small files, especially smaller than the default block size.
  • Hadoop cannot be used for running interactive jobs or analytics.
Read full review
  • Teradata can improve by supporting more native AWS cloud features. Currently if a node goes down the EC2 instance must be restarted. It isn't something that happens frequently but more tight integration with cloud providers like AWS and Azure will allow Teradata to offer truly dynamic scaling.
  • Some Teradata features are oversold before they are ready for prime-time. Teradata is not unique in this but if something is sold as an integrated product stack it should really be integrated not something that requires an extensive development cycle to be integrated at a customer's expense. If something is supported it should've really be tested and QAed thoroughly before a customer touches it.
Read full review
Likelihood to Renew
Hadoop is organization-independent and can be used for various purposes ranging from archiving to reporting and can make use of economic, commodity hardware. There is also a lot of saving in terms of licensing costs - since most of the Hadoop ecosystem is available as open-source and is free
Read full review
Teradata is a mature RDBMS system that expands its functionality towards the current cloud capabilities like object storage and flexible compute scale.
Read full review
Usability
Great! Hadoop has an easy to use interface that mimics most other data warehouses. You can access your data via SQL and have it display in a terminal before exporting it to your business intelligence platform of choice. Of course, for smaller data sets, you can also export it to Microsoft Excel.
Read full review
Teradata Vantage allows us to create a scalable infrastructure to support our strategic initiatives. The dedicated compute power ensures reliable performance with isolated workloads and dedicated resources, optimizing workflows for faster, more efficient data transfers. The compute clusters support ETL processes and OSF’s developers and data science team with the flexibility to create self-service analytics, to spin up/down at any time, driving better performance and minimizing costs.
Read full review
Support Rating
We went with a third party for support, i.e., consultant. Had we gone with Azure or Cloudera, we would have obtained support directly from the vendor. my rating is more on the third party we selected and doesn't reflect the overall support available for Hadoop. I think we could have done better in our selection process, however, we were trying to use an already approved vendor within our organization. There is plenty of self-help available for Hadoop online.
Read full review
We have meetings at the beginning with the technical team to explain our requirements to them and they were really putting in a lot of effort to come up with a solution which will address all our needs. They implemented the software and also trained a few of our resources on the same too. We can get in touch with them now as well whenever we run into a roadblock but it's very less now.
Read full review
Online Training
Hadoop is a complex topic and best suited for classrom training. Online training are a waste of time and money.
Read full review
No answers on this topic
Alternatives Considered
I feel that this is a highly reliable and scalable solution computing technology that is highly capable of processing large data sets across multiple servers and thousands of machines in a well-defined and distributed manner. Apache Hadoop can automatically scale up the number of servers and machines that are needed to process, store, and analyze data sets. It also handles explosions in data with big data technology. Apache Hadoop is good at handling all node failures as well.
Read full review
Teradata is way ahead of its competitor because of its unique features of ensuring data privacy and data never gets corrupted even in worst case scenario. In most cases, the data corruption is a major issue if left unused and it leads to important data being wiped off which in ideal case should be stored for 3 years
Read full review
Return on Investment
  • As it was open source makes it popular choice for handling large chuck of datasets
  • It was free earlier but now it’s licensed but still enterprise is a fine tuned version which makes it easier for new users and administrators to use it
  • Our investment is worth every single penny.
  • Initial cost is more as you might need to hire administrators to setup the cluster and make them in scalable. But once done it’s pretty easy
Read full review
  • Teradata is been absolutely phenomenal for our project because we feed huge chunks of data to it and get back the desired results in no time which earlier used to take hours to process and then also sometimes timeout.
  • We don't have to do any manual intervention for resource or task allocation, it is all taken care by Teradata internally and all the AMP's are given equal amount of work and have their own resources to complete them with no sharing with another.
Read full review
ScreenShots

Teradata Vantage Screenshots

Screenshot of Teradata VantageCloud Lake Console Financial GovernanceScreenshot of Teradata VantageCloud Lake Console Landing Page