Cookies help us display personalized product recommendations and ensure you have great shopping experience.

By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
SmartData CollectiveSmartData Collective
  • Analytics
    AnalyticsShow More
    New Data Analytics Breakthroughs Give eCommerce Startups a Fighting Chance
    New Data Analytics Breakthroughs Give eCommerce Startups a Fighting Chance
    6 Min Read
    How Data Analytics Is Reshaping Patient Financing Decisions
    How Data Analytics Is Reshaping Patient Financing Decisions
    13 Min Read
    business using business intelligence
    How to Use a Competitive Intelligence Dashboard to Turn Market Data Into Smarter Marketing Decisions 
    9 Min Read
    unusual trading activity
    Signal Or Noise? A Decision Tree For Evaluating Unusual Trading Activity
    3 Min Read
    software developer using ai
    How Data Analytics Helps Developers Deliver Better Tech Services
    8 Min Read
  • Big Data
  • BI
  • Exclusive
  • IT
  • Marketing
  • Software
Search
© 2008-25 SmartData Collective. All Rights Reserved.
Reading: 3 Big Hadoop Myths Dispelled
Share
Notification
Font ResizerAa
SmartData CollectiveSmartData Collective
Font ResizerAa
Search
  • About
  • Help
  • Privacy
Follow US
© 2008-23 SmartData Collective. All Rights Reserved.
SmartData Collective > Big Data > Data Warehousing > 3 Big Hadoop Myths Dispelled
Big DataData WarehousingHadoopITMapReduceOpen SourceSoftwareSQL

3 Big Hadoop Myths Dispelled

MicheleNemschoff
MicheleNemschoff
4 Min Read
Image
SHARE

ImageAs with any new technological innovation, a lot of myths have been generated about Hadoop. At Strata + Hadoop World, held in NY at the end of October, Jack Norris, CMO of MapR, discussed three of those myths and how MapR dispels them.

 1. Hadoop Distributions are Incredibly Competitive

ImageAs with any new technological innovation, a lot of myths have been generated about Hadoop. At Strata + Hadoop World, held in NY at the end of October, Jack Norris, CMO of MapR, discussed three of those myths and how MapR dispels them.

 1. Hadoop Distributions are Incredibly Competitive

More Read

User Adoption
User Adoption – Resistance Is Futile, We Hope
Differentiating Between Data Lakes and Data Warehouses
A Particularly Snarky Interview with Joe Celko
How To Use Big Data As Part Of Your Investment Planning
Getting business value from data? Commercial analytics is where it’s at

Norris declared the vicious competition between Hadoop distributions a myth. There are many commercial Hadoop platforms to choose from, but they all share the same open source code. Since Hadoop was created by open source technology and is in its early stages, it has been necessary for commercial distributors to add their own innovations to the open source code to meet the needs of the enterprise customer. The type of innovations have varied, with some adding management functions on top of the open source product and others adapting the underlying architecture. As a result, there is a diverse ecosystem of Hadoop distributions available that, as the fastest growing big data technology, has been a huge job creator.

2. All NoSQL Solutions are Created Equal

Apache HBase, a NoSQL solution, is integrated with Hadoop and included in every commercial Hadoop distribution. However, just because HBase is included in each distribution does not not mean that it runs the same in each distribution because the architecture that supports HBase varies. A typical architecture for HBase includes multiple layers. HBase will run on Java which runs on HDFS which writes into the Linux file system which writes to the disk. Distributions will vary on how many layers are required to write HDFS, and the more complex those layers are the worse the performance. On the other hand, if those layers are condensed, performance is greatly improved.

3. Hadoop isn’t Enterprise Ready

The most popular criticism of Hadoop is that it isn’t ready to provide real value to enterprises. However, the number of businesses currently having success with Hadoop are more than enough to prove this theory is a myth. 

For example, Solutionary analyzes and processes 1 trillion log lines for its security service, and Rubicon Project processes 90 billion ad auctions per day with Hadoop. If you are looking for examples of companies that aren’t Web 2.0, consider a Fortune 100 retailer that runs more than 2000 nodes on Hadoop to use social media to better reach out to its customers or a financial services company that uses Hadoop to mitigate risk and create personalized offers for its customers.

The reality is Hadoop is being used in a variety of industries from Web 2.0 to waste management and healthcare. Currently, the Climate Corporation uses Hadoop to help farmers improve their crop production, and a beverage company in Japan has created kiosks that use facial recognition to create a customized interface.

Hadoop is an incredibly valuable technology that unfortunately is misunderstood by many. While Hadoop distributions vary, especially in their integration of NoSQL, there are enterprise-ready options available that many enterprises are already having success with.

Share This Article
Facebook Pinterest LinkedIn
Share

Follow us on Facebook

Latest News

Why Every Small Business Should Care About an AI Image Generator
Why Every Small Business Should Care About an AI Image Generator
Artificial Intelligence Exclusive
ai for instagram reel marketing
How AI Is Changing Instagram Reel Marketing
Artificial Intelligence Exclusive Marketing
protecting data in public
The Importance Of Protecting Sensitive Data In Public Services
Big Data Data Management Exclusive
New Data Analytics Breakthroughs Give eCommerce Startups a Fighting Chance
New Data Analytics Breakthroughs Give eCommerce Startups a Fighting Chance
Analytics Big Data Exclusive

Stay Connected

1.2KFollowersLike
33.7KFollowersFollow
222FollowersPin

You Might also Like

Image
Hardware

Overcoming the Wearable Tech Battery Life Challenge

6 Min Read
google+ and big data analytics
AnalyticsBig DataExclusiveSocial DataSocial Media Analytics

Google+ Is After Your Friends with Big Data and Beautiful Photos

5 Min Read
unstructured data
AnalyticsBest PracticesBig DataBusiness IntelligenceCloud ComputingData ManagementITMarketingMobilitySocial DataSocial Media AnalyticsUnstructured Data

Managing Unstructured Data: The Next BI Point of Emphasis

3 Min Read
startups and IT
Cloud ComputingHardwareITSoftware

The Start-ups Guide to IT: Where to Invest your Money for Business Technology

3 Min Read

SmartData Collective is one of the largest & trusted community covering technical content about Big Data, BI, Cloud, Analytics, Artificial Intelligence, IoT & more.

data-driven web design
5 Great Tips for Using Data Analytics for Website UX
Big Data
giveaway chatbots
How To Get An Award Winning Giveaway Bot
Big Data Chatbots Exclusive

Quick Link

  • About
  • Contact
  • Privacy
Follow US
© 2008-25 SmartData Collective. All Rights Reserved.
Welcome Back!

Sign in to your account

Username or Email Address
Password

Lost your password?