By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
SmartData CollectiveSmartData Collective
  • Analytics
    AnalyticsShow More
    AI analytics
    AI-Based Analytics Are Changing the Future of Credit Cards
    6 Min Read
    data overload showing data analytics
    How Does Next-Gen SIEM Prevent Data Overload For Security Analysts?
    8 Min Read
    hire a marketing agency with a background in data analytics
    5 Reasons to Hire a Marketing Agency that Knows Data Analytics
    7 Min Read
    predictive analytics for amazon pricing
    Using Predictive Analytics to Get the Best Deals on Amazon
    8 Min Read
    data science anayst
    Growing Demand for Data Science & Data Analyst Roles
    6 Min Read
  • Big Data
  • BI
  • Exclusive
  • IT
  • Marketing
  • Software
Search
© 2008-23 SmartData Collective. All Rights Reserved.
Reading: 3 Big Hadoop Myths Dispelled
Share
Notification Show More
Aa
SmartData CollectiveSmartData Collective
Aa
Search
  • About
  • Help
  • Privacy
Follow US
© 2008-23 SmartData Collective. All Rights Reserved.
SmartData Collective > Big Data > Data Warehousing > 3 Big Hadoop Myths Dispelled
Big DataData WarehousingHadoopITMapReduceOpen SourceSoftwareSQL

3 Big Hadoop Myths Dispelled

MicheleNemschoff
Last updated: 2014/03/04 at 8:19 PM
MicheleNemschoff
4 Min Read
Image
SHARE

ImageAs with any new technological innovation, a lot of myths have been generated about Hadoop. At Strata + Hadoop World, held in NY at the end of October, Jack Norris, CMO of MapR, discussed three of those myths and how MapR dispels them.

 1. Hadoop Distributions are Incredibly Competitive

ImageAs with any new technological innovation, a lot of myths have been generated about Hadoop. At Strata + Hadoop World, held in NY at the end of October, Jack Norris, CMO of MapR, discussed three of those myths and how MapR dispels them.

 1. Hadoop Distributions are Incredibly Competitive

More Read

Data Ethics: Safeguarding Privacy and Ensuring Responsible Data Practices

Data Ethics: Safeguarding Privacy and Ensuring Responsible Data Practices

8 Crucial Tips to Help SMEs Guard Against Data Breaches
Banks Merge Data Mining and CRM Tools to Boost Profitability
Cloud Advances Make Record Keeping Compliance Easier Than Ever
Digital Transformation: How To Protect Your Organization From Cyber Risk

Norris declared the vicious competition between Hadoop distributions a myth. There are many commercial Hadoop platforms to choose from, but they all share the same open source code. Since Hadoop was created by open source technology and is in its early stages, it has been necessary for commercial distributors to add their own innovations to the open source code to meet the needs of the enterprise customer. The type of innovations have varied, with some adding management functions on top of the open source product and others adapting the underlying architecture. As a result, there is a diverse ecosystem of Hadoop distributions available that, as the fastest growing big data technology, has been a huge job creator.

2. All NoSQL Solutions are Created Equal

Apache HBase, a NoSQL solution, is integrated with Hadoop and included in every commercial Hadoop distribution. However, just because HBase is included in each distribution does not not mean that it runs the same in each distribution because the architecture that supports HBase varies. A typical architecture for HBase includes multiple layers. HBase will run on Java which runs on HDFS which writes into the Linux file system which writes to the disk. Distributions will vary on how many layers are required to write HDFS, and the more complex those layers are the worse the performance. On the other hand, if those layers are condensed, performance is greatly improved.

3. Hadoop isn’t Enterprise Ready

The most popular criticism of Hadoop is that it isn’t ready to provide real value to enterprises. However, the number of businesses currently having success with Hadoop are more than enough to prove this theory is a myth. 

For example, Solutionary analyzes and processes 1 trillion log lines for its security service, and Rubicon Project processes 90 billion ad auctions per day with Hadoop. If you are looking for examples of companies that aren’t Web 2.0, consider a Fortune 100 retailer that runs more than 2000 nodes on Hadoop to use social media to better reach out to its customers or a financial services company that uses Hadoop to mitigate risk and create personalized offers for its customers.

The reality is Hadoop is being used in a variety of industries from Web 2.0 to waste management and healthcare. Currently, the Climate Corporation uses Hadoop to help farmers improve their crop production, and a beverage company in Japan has created kiosks that use facial recognition to create a customized interface.

Hadoop is an incredibly valuable technology that unfortunately is misunderstood by many. While Hadoop distributions vary, especially in their integration of NoSQL, there are enterprise-ready options available that many enterprises are already having success with.

MicheleNemschoff March 4, 2014
Share This Article
Facebook Twitter Pinterest LinkedIn
Share

Follow us on Facebook

Latest News

Data Ethics: Safeguarding Privacy and Ensuring Responsible Data Practices
Data Ethics: Safeguarding Privacy and Ensuring Responsible Data Practices
Best Practices Big Data Data Collection Data Management Privacy
data protection for SMEs
8 Crucial Tips to Help SMEs Guard Against Data Breaches
Data Management
How AI is Boosting the Customer Support Game
How AI is Boosting the Customer Support Game
Artificial Intelligence
AI analytics
AI-Based Analytics Are Changing the Future of Credit Cards
Analytics Artificial Intelligence Exclusive

Stay Connected

1.2k Followers Like
33.7k Followers Follow
222 Followers Pin

You Might also Like

Data Ethics: Safeguarding Privacy and Ensuring Responsible Data Practices
Best PracticesBig DataData CollectionData ManagementPrivacy

Data Ethics: Safeguarding Privacy and Ensuring Responsible Data Practices

7 Min Read
data protection for SMEs
Data Management

8 Crucial Tips to Help SMEs Guard Against Data Breaches

10 Min Read
data mining and crm for banking
Big Data

Banks Merge Data Mining and CRM Tools to Boost Profitability

9 Min Read
cloud advances
Cloud Computing

Cloud Advances Make Record Keeping Compliance Easier Than Ever

8 Min Read

SmartData Collective is one of the largest & trusted community covering technical content about Big Data, BI, Cloud, Analytics, Artificial Intelligence, IoT & more.

AI and chatbots
Chatbots and SEO: How Can Chatbots Improve Your SEO Ranking?
Artificial Intelligence Chatbots Exclusive
ai in ecommerce
Artificial Intelligence for eCommerce: A Closer Look
Artificial Intelligence

Quick Link

  • About
  • Contact
  • Privacy
Follow US
© 2008-23 SmartData Collective. All Rights Reserved.
Go to mobile version
Welcome Back!

Sign in to your account

Lost your password?