We use cookies, including third-party cookies from Google to serve personalized ads through AdSense, to operate this site and understand how it is used. By continuing to browse, you accept this use. See our Privacy Policy and Terms of Use for details, including how to opt out of personalized advertising.
Accept
SmartData CollectiveSmartData Collective
  • Analytics
    AnalyticsShow More
    What Kind of Problem-Solving Distinguishes Data Analysts From Software Engineers -- AI-generated illustration
    What Kind of Problem-Solving Distinguishes Data Analysts From Software Engineers
    7 Min Read
    chatgpt image jul 21, 2026, 04 34 30 pm
    4 Core Benefits of Predictive Maintenance after Vibration Analysis
    10 Min Read
    How Does Data Mining Boost Customer Satisfaction in Logistics? Harnessing Analytics for Results -- AI-generated illustration
    How Does Data Mining Boost Customer Satisfaction in Logistics? Harnessing Analytics for Results
    11 Min Read
    chatgpt image jul 13, 2026, 04 23 45 pm
    How Data Analytics Helps Companies Improve User Engagement
    19 Min Read
    chatgpt image jul 13, 2026, 03 59 46 pm
    How Data Analytics Improves Multi-Location Search Strategies
    10 Min Read
  • Big Data
  • BI
  • Exclusive
  • IT
  • Marketing
  • Software
Search
© 2008-25 SmartData Collective. All Rights Reserved.
Reading: Who knows what happiness lurks in the hearts of men? Facebook knows.
Share
Notification
Font ResizerAa
SmartData CollectiveSmartData Collective
Font ResizerAa
Search
  • About
  • Help
  • Privacy
Follow US
© 2008-23 SmartData Collective. All Rights Reserved.
SmartData Collective > Big Data > Data Mining > Who knows what happiness lurks in the hearts of men? Facebook knows.
Data Mining

Who knows what happiness lurks in the hearts of men? Facebook knows.

DavidBakken
DavidBakken
8 Min Read
Who knows what happiness lurks in the hearts of men?  Facebook knows.
Illustration generated with FLUX.2 [klein 4B] via Cloudflare Workers AI.
SHARE

Have you heard about the Facebook Gross National Happiness Index?  On Monday, October 12, the Times ran an article (by Noam Cohen) reporting some of the findings based on analysis of two years’ worth of Facebook status updates from 100 million users in the U.S. The index was created by Adam D. I. Kramer, a doctoral candidate in social psychology at the University of Oregon, and is based on counts of positive and negative words in status updates. According to the article, classification of words as positive or negative is based on the Linguistic Inquiry and Word Count dictionary.

Among the researchers’ conclusions: we’re happier on Fridays than on Mondays; holidays also make Americans happy. The premature death of a celebrity may make us sad. According to a post by Mr. Kramer on the Facebook blog, the two “saddest” days – days with the highest numbers of negative words – were the days on which actor Heath Ledger and pop icon Michael Jackson died. Mr. Kramer points out that, coincidentally, Mr. Ledger died on the day of the Asian stock market crash, which might have contributed to the degree of negativity.

We’re going to see a lot more of this kind of thing as researchers delve into the rich trove of information generated by users of search engines and web-enabled social networking. The happiness index, based as it is on simple frequency analysis of words, is the tip of the iceberg. At the moment, “social media” – I’m not exactly sure what that label means – is getting incredible attention in the marketing and marketing research community. The question that has yet to be posed, let alone answered, is, “what exactly do we learn from all this information?”

More Read

A Look at Today’s White House Big Data Event
A Look at Today’s White House Big Data Event
Can Big Data and Hadoop Feed the World?
Using Cell Phone Data for Social Good
Guest Post: Can Database Developers do Data Mining ?
Revolution Analytics Hosts Contest on Business Predicting the Future

The Facebook Gross Happiness Index is revealing. We can see a pattern in the data. Status updates contain more words that are positive on occasions when we might expect greater positive sentiment, and fewer positive words on those days when we might expect greater negative sentiment. But anytime we look at spontaneously generated data, we need to ask “what’s missing?” User-generated content on the web is subject to coverage and selection biases which can sometimes be quite large. The happiness index provides an example. Is the pronounced negative sentiment on the day Heath Ledger died a function of the demographics of the Facebook community, or the subset who chose to update their status on that day, or some other unidentified subset of facebook users?

In the “what’s old is new” category, a book published more than forty years ago can serve a guide to using social media and user-generated content for research. Unobtrusive Measures: Nonreactive Research in the Social Sciences shed much light on the problems of traditional surveys, and offered a structured approach to using, in a sense, “found” data for social research.  This book was out of print for a long time but happily has been reissued (at about 20 times the price I paid when I bought my copy as an undergraduate). While this book was out of print, Raymond M. Lee published Unobtrusive Measures in Social Research, which incorporated and updated the ideas in the original. Lee’s book includes sections on using the Internet for social science research. Either of these books should be required reading for anyone attempting to use social media and user-generated content for research purposes. The original Unobtrusive Measures exemplifies the best in social science writing, and is a pleasure to read.

The key to success with data generated from online activity, as with all research, is understanding the limitations in the data source and finding ways to compensate for those limitations. The authors of the original Unobtrusive Measures (Eugene J. Webb, Donald T. Campbell, Richard D. Schwartz, and Lee Sechrest) argue in favor of using multiple approaches. They view social science research as an “approximation” to knowledge, and the more points or manifestations we observe for each phenomenon of interest, the closer our approximation gets to some underlying “truth.”

The type of analysis reflected in the Facebook index has many limitations. Given the large number of individuals providing data, it’s tempting to discount potential biases due to coverage and selection. There are many potential individual-level causes that get rolled up into the aggregate numbers of positive and negative words in the updates posted on any given day. We might assume that Facebook has the potential to disaggregate the data to a degree, and provide more insight into the factors that might drive swings in the happiness index. That’s something we should look forward to seeing from them.

Copyright 2009 by David G. Bakken.  All rights reserved.

Share This Article
Facebook Pinterest LinkedIn
Share

Follow us on Facebook

Latest News

What Kind of Problem-Solving Distinguishes Data Analysts From Software Engineers -- AI-generated illustration
What Kind of Problem-Solving Distinguishes Data Analysts From Software Engineers
Analytics Big Data Exclusive Software
Flat editorial illustration: The article examines AI agents that escalate from legitimate data retrieval to attempted intrusions
OpenAI’s Government Website Incidents Raise a Hard Question for AI Agents: When Should They Stop?
Artificial Intelligence News Security
Flat editorial illustration: The article's core relationship is the alignment between customer behavioral data (visit frequency,
Data-Driven Loyalty: How Restaurants Use Behavioral Analytics to Optimize Revenue
Exclusive
Flat editorial illustration: The article's core relationship is that reliable eCommerce attribution depends on a unified, well-st
How eCommerce Data Teams Can Build Attribution That Holds Up
Big Data Exclusive

Stay Connected

1.2KFollowersLike
33.7KFollowersFollow
222FollowersPin

You Might also Like

Truly Distributed Analytics
Data Mining

Truly Distributed Analytics

4 Min Read
Big Data Analytics Versus Your Own Lying Eyes
AnalyticsData MiningPredictive AnalyticsStatistics

Big Data Analytics Versus Your Own Lying Eyes

0 Min Read
Using Sentiment to Understand Your Consumer & Your Competitors
AnalyticsData MiningData Visualization

Using Sentiment to Understand Your Consumer & Your Competitors

4 Min Read
Text Analytics, Big Data and the Keys to ROI
AnalyticsBest PracticesData MiningHadoopText AnalyticsUnstructured Data

Text Analytics, Big Data and the Keys to ROI

11 Min Read

SmartData Collective is one of the largest & trusted community covering technical content about Big Data, BI, Cloud, Analytics, Artificial Intelligence, IoT & more.

Artificial Intelligence for eCommerce: A Closer Look
Artificial Intelligence for eCommerce: A Closer Look
Artificial Intelligence
AI chatbots
AI Chatbots Can Help Retailers Convert Live Broadcast Viewers into Sales!
Chatbots

Quick Link

  • About
  • Contact
  • Privacy
Follow US
© 2008-26 SmartData Collective. All Rights Reserved.
Welcome Back!

Sign in to your account

Username or Email Address
Password

Lost your password?