We use cookies, including third-party cookies from Google to serve personalized ads through AdSense, to operate this site and understand how it is used. By continuing to browse, you accept this use. See our Privacy Policy and Terms of Use for details, including how to opt out of personalized advertising.
Accept
SmartData CollectiveSmartData Collective
  • Analytics
    AnalyticsShow More
    chatgpt image jul 21, 2026, 04 34 30 pm
    4 Core Benefits of Predictive Maintenance after Vibration Analysis
    10 Min Read
    How Does Data Mining Boost Customer Satisfaction in Logistics? Harnessing Analytics for Results -- AI-generated illustration
    How Does Data Mining Boost Customer Satisfaction in Logistics? Harnessing Analytics for Results
    11 Min Read
    chatgpt image jul 13, 2026, 04 23 45 pm
    How Data Analytics Helps Companies Improve User Engagement
    19 Min Read
    chatgpt image jul 13, 2026, 03 59 46 pm
    How Data Analytics Improves Multi-Location Search Strategies
    10 Min Read
    cybersecurity efforts
    How Behavioral Analytics and AI Are Redefining Cybersecurity for Boca Raton Businesses
    14 Min Read
  • Big Data
  • BI
  • Exclusive
  • IT
  • Marketing
  • Software
Search
© 2008-25 SmartData Collective. All Rights Reserved.
Reading: What Is Your Big Data Analytics Stack?
Share
Notification
Font ResizerAa
SmartData CollectiveSmartData Collective
Font ResizerAa
Search
  • About
  • Help
  • Privacy
Follow US
© 2008-23 SmartData Collective. All Rights Reserved.
SmartData Collective > Analytics > Predictive Analytics > What Is Your Big Data Analytics Stack?
Predictive Analytics

What Is Your Big Data Analytics Stack?

Radhika Subramanian
Radhika Subramanian
5 Min Read
What Is Your Big Data Analytics Stack?
Illustration generated with FLUX.2 [klein 4B] via Cloudflare Workers AI.
SHARE

We often get asked this question – Where do I begin?  How are problems being solved using big-data analytics?

To answer this question we need to take a step back and think in the context of the problem and a complete solution to the problem.

The objective of big data, or any data for that matter, is to solve a business problem. The business problem is also called a use-case. We always keep that in mind. The easiest way to explain the data stack is by starting at the bottom, even though the process of building the use-case is from the top.

We often get asked this question – Where do I begin?  How are problems being solved using big-data analytics?

More Read

Predictive Analytics: The Power and the Gory
Predictive Analytics: The Power and the Gory
The New Predictive Profession: Odd Yet Newly Legitimate [BOOK REVIEW]
Access Layer Data and Sensor-2-Server
Big Data, Unstructured Information Analysis is More Than Sentiment.
Emotions: The Next (but not new) Frontier in Artificial Intelligence & Cognitive Computing

To answer this question we need to take a step back and think in the context of the problem and a complete solution to the problem.

The objective of big data, or any data for that matter, is to solve a business problem. The business problem is also called a use-case. We always keep that in mind. The easiest way to explain the data stack is by starting at the bottom, even though the process of building the use-case is from the top.

Data Layer: The bottom layer of the stack, of course, is data. This is the raw ingredient that feeds the stack. The players here are the database and storage vendors. Hadoop, with its innovative approach, is making a lot of waves in this layer.

Data Preparation Layer: The next layer is the data preparation tool. As we all know, data is typically messy and never in the right form. Data preparation is the process of extracting data from the source(s), merging two data sets and preparing the data required for the analysis step. There are emerging players in this area. 

Analysis Layer: The next layer is the analysis layer. Statistics is the most commonly known analysis tool. For statistics, the commonly available solutions are statistics and open source R. This is the layer for the emerging machine learning solutions. Automated analysis with machine learning is the future.  

Presentation Layer: The output from the analysis engine feeds the presentation layer. The presentation layer depends on the use-case. This layer is called the action layer, consumption layer or last mile.

  • If the result of the use case is to be presented to a human, the presentation layer may be a BI or visualization tool.   Example use-cases are fraud detection, Order-to-cash monitoring, etc. In each case the final result is sent to human decision makers for them to act.
  • For some use-cases, the results need to feed a downstream system, which may be another program. Example use-cases are recommendation systems, real-time pricing systems, etc. In this case the analysis results are fed into the downstream system that acts on it.
  • If the use-case is an alerting system, then the analysis results feed an event processing or alerting system. Example use-cases are medical device failure, network failure, etc. In this case the results of the analysis are fed into a system that can send out alerts to humans or machines that will act on the results in real-time or near real-time.

Use-case Layer: This is the value layer, and the ultimate purpose of the entire data stack. The use-case drives the selection of tools in each layer of the data stack. The number of use-cases is practically infinite. Example use-cases are fraud detection, dropped call alerting, network failure, supplier failure alerting, machine failure, and so on. These are like recipes in cookbooks – practically infinite. As the types and amount of data grows, the number of use-cases will grow.

How do you think about your data stack?

What are your thoughts?

 

Share This Article
Facebook Pinterest LinkedIn
Share

Follow us on Facebook

Latest News

Flat editorial illustration: The article's core relationship is the brand protection response workflow: detection of a phishing o
Data & AI Architecture Focus: 6 Best Brand Protection Tools for Phishing and Impersonation
IT Security
Server racks with cloud and user interface panels
Cloud Infrastructure and Workload Migration: A Data-Driven Look at VMware Alternatives in Europe
Cloud Computing Exclusive
Synthetic Data vs Real Web Data: Comparison, Limitations, and Collection Methods  -- AI-generated illustration
Synthetic Data vs Real Web Data: Comparison, Limitations, and Collection Methods 
Big Data Exclusive
Illustration of mobile analytics dashboards with ad performance charts connected to backend databases
11 Best Sisense Alternatives for Embedded Analytics
Business Intelligence Exclusive

Stay Connected

1.2KFollowersLike
33.7KFollowersFollow
222FollowersPin

You Might also Like

Do Predictive Modelers Need to Know Math?
AnalyticsModelingPredictive Analytics

Do Predictive Modelers Need to Know Math?

6 Min Read
Predictive Analytics in Action: Anthony Goldbloom of Kaggle
Big DataInside CompaniesModelingPredictive Analytics

Predictive Analytics in Action: Anthony Goldbloom of Kaggle

6 Min Read
The New Way to Segment For a 6x Greater Return
AnalyticsBest PracticesData MiningMarketingPredictive Analytics

The New Way to Segment For a 6x Greater Return

5 Min Read
What About the Rest of Us?
Business IntelligenceData MiningPredictive Analytics

What About the Rest of Us?

4 Min Read

SmartData Collective is one of the largest & trusted community covering technical content about Big Data, BI, Cloud, Analytics, Artificial Intelligence, IoT & more.

Chatbots and SEO: How Can Chatbots Improve Your SEO Ranking?
Chatbots and SEO: How Can Chatbots Improve Your SEO Ranking?
Artificial Intelligence Chatbots Exclusive
How To Get An Award Winning Giveaway Bot
How To Get An Award Winning Giveaway Bot
Big Data Chatbots Exclusive

Quick Link

  • About
  • Contact
  • Privacy
Follow US
© 2008-26 SmartData Collective. All Rights Reserved.
Welcome Back!

Sign in to your account

Username or Email Address
Password

Lost your password?