By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
SmartData CollectiveSmartData CollectiveSmartData Collective
  • Analytics
    AnalyticsShow More
    data-driven white label SEO
    Does Data Mining Really Help with White Label SEO?
    7 Min Read
    marketing analytics for hardware vendors
    IT Hardware Startups Turn to Data Analytics for Market Research
    9 Min Read
    big data and digital signage
    The Power of Big Data and Analytics in Digital Signage
    5 Min Read
    data analytics investing
    Data Analytics Boosts ROI of Investment Trusts
    9 Min Read
    football data collection and analytics
    Unleashing Victory: How Data Collection Is Revolutionizing Football Performance Analysis!
    4 Min Read
  • Big Data
  • BI
  • Exclusive
  • IT
  • Marketing
  • Software
Search
© 2008-23 SmartData Collective. All Rights Reserved.
Reading: Hygienic Hadoop Data Lakes Not Just Happenstance
Share
Notification Show More
Aa
SmartData CollectiveSmartData Collective
Aa
Search
  • About
  • Help
  • Privacy
Follow US
© 2008-23 SmartData Collective. All Rights Reserved.
SmartData Collective > Software > Hadoop > Hygienic Hadoop Data Lakes Not Just Happenstance
CommentaryData ManagementExclusiveHadoopOpen SourcePolicy and Governance

Hygienic Hadoop Data Lakes Not Just Happenstance

paulbarsch
Last updated: 2015/02/23 at 10:52 AM
paulbarsch
4 Min Read
Image
SHARE

Image

Image

It is often thought that Apache® Hadoop based data lakes are a potential panacea to thorny data management issues long germane to relational databases. After all, the (mistaken) belief goes, you can simply dump all your data into Hadoop’s file system, and via schema on read magic, your desired result sets will appear with very little effort. However, data management—even for Hadoop—isn’t going away and in fact, probably never will.

If you’ve never read Tyler Brûlé’s columns in the Financial Times, you’re really missing something. Mr. Brûlé’s column is a Sunday morning staple where he comments on design, style, business, travel and more.  Even better, Mr. Brûlé was recently paired with the FT’s Lucy Kellaway in an article where they discussed Mr. Brûlé’s obsession with cleanliness, order, and aesthetics.

More Read

Image

4 Business Risks That Might Prevent Big Data ROI

3 Big Data Potholes to Avoid
Is Cloud Sameness Dangerous to Competitive Advantage?
Wasted Breath: Data Alone Won’t Convince
Less Dogma Equals Better Decision Making

Reading along, I found some interesting parallels with Mr. Brûlé’s observations on office clutter, and relational database design practices.

For example, Mr. Brûlé despises anarchy. In the article he pointed to a staffer’s empty Evian water bottle on a desk, complaining that such items take away from his emphasis on office décor. And most certainly, Mr. Brûlé does not like jackets on the back of chairs, and anything else that takes away from the intended design and decoration of the office.  Why? “There needs to be a rule of law, or else where does it end?” he says. Otherwise “people will come in with wheelie suitcases, or with plastic hangers and dry cleaning.”

While some may find all this attention to detail slightly amusing, Mr. Brûlé does not. And neither does your company database administrator (DBA). That’s because in order to provide accurate reports, BI visualizations and powerful analytics, there are significant efforts that must take place to identify, model, transform, curate and secure data in a relational database. In essence, all the work to make your data look as clean, ordered and useful as Mr. Brûlé’s office is an ongoing process handled by your data stewards and DBAs.

Now let’s get back to Hadoop. Almost no one who works with Hadoop on a daily basis would suggest that data can simply be dumped into Hadoop’s file system and be of high value to rank and file business users.

Want to store sensitive data in your data lake? You’ll most certainly be doubling down on your efforts to lockup key data, especially since Hadoop security is evolving. In addition to data security, you’ll still have to contend with metadata management, architecture and design, and governance in Hadoop. Indeed, none of these data management issues are going away if you’re planning on allowing Hadoop to serve as a true lake or “hub” for all your organization’s data.

Are you just starting out with your Hadoop data lake and not quite there yet in terms of clear and accepted data management processes? It probably seems like a gargantuan task at first, but as the old yarn goes, you eat the elephant one bite at time. Or in the words of the esteemed Tyler Brûlé; “It’s a daily effort to adhere to set standards. (But) you need to aspire to something.”  Even if that “something” is a well governed, managed and secured Hadoop based data lake.

TAGGED: risky business
paulbarsch February 23, 2015
Share This Article
Facebook Twitter Pinterest LinkedIn
Share

Follow us on Facebook

Latest News

big data and IP laws
Big Data & AI In Collision Course With IP Laws – A Complete Guide
Big Data
ai in marketing
4 Ways AI Can Enhance Your Marketing Strategies
Marketing
sobm for ai-driven cybersecurity
Software Bill of Materials is Crucial for AI-Driven Cybersecurity
Security
IT budgeting for data-driven companies
IT Budgeting Practices for Data-Driven Companies
IT

Stay Connected

1.2k Followers Like
33.7k Followers Follow
222 Followers Pin

You Might also Like

Image
Big DataBusiness IntelligenceRisk Management

4 Business Risks That Might Prevent Big Data ROI

5 Min Read
Image
Big Data

3 Big Data Potholes to Avoid

5 Min Read
Image
Uncategorized

Is Cloud Sameness Dangerous to Competitive Advantage?

5 Min Read
Image
CommentaryExclusive

Wasted Breath: Data Alone Won’t Convince

5 Min Read

SmartData Collective is one of the largest & trusted community covering technical content about Big Data, BI, Cloud, Analytics, Artificial Intelligence, IoT & more.

ai is improving the safety of cars
From Bolts to Bots: How AI Is Fortifying the Automotive Industry
Artificial Intelligence
data-driven web design
5 Great Tips for Using Data Analytics for Website UX
Big Data

Quick Link

  • About
  • Contact
  • Privacy
Follow US
© 2008-23 SmartData Collective. All Rights Reserved.
Go to mobile version
Welcome Back!

Sign in to your account

Lost your password?