Cookies help us display personalized product recommendations and ensure you have great shopping experience.

By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
SmartData CollectiveSmartData Collective
  • Analytics
    AnalyticsShow More
    warehouse accidents
    Data Analytics and the Future of Warehouse Safety
    10 Min Read
    stock investing and data analytics
    How Data Analytics Supports Smarter Stock Trading Strategies
    4 Min Read
    predictive analytics risk management
    How Predictive Analytics Is Redefining Risk Management Across Industries
    7 Min Read
    data analytics and gold trading
    Data Analytics and the New Era of Gold Trading
    9 Min Read
    composable analytics
    How Composable Analytics Unlocks Modular Agility for Data Teams
    9 Min Read
  • Big Data
  • BI
  • Exclusive
  • IT
  • Marketing
  • Software
Search
© 2008-25 SmartData Collective. All Rights Reserved.
Reading: Voodoo Spectrum of Machine Learning and Data Sets
Share
Notification
Font ResizerAa
SmartData CollectiveSmartData Collective
Font ResizerAa
Search
  • About
  • Help
  • Privacy
Follow US
© 2008-23 SmartData Collective. All Rights Reserved.
SmartData Collective > Big Data > Data Mining > Voodoo Spectrum of Machine Learning and Data Sets
Business IntelligenceData Mining

Voodoo Spectrum of Machine Learning and Data Sets

Editor SDC
Editor SDC
3 Min Read
SHARE

I used to be very gung-ho about machine learning approaches to trading but I’m less so now. You have to understand that that there is a spectrum of alpha sources, from very specific structured arbitrage opportunities -> to stat arb -> to just voodoo nonsense.

As history goes on, hedge funds and other large players are absorbing the alpha from left to right. Having squeezed the pure arbs (ADR vs underlying, ETF vs components, mergers, currency triangles, etc) they then became hungry again and moved to stat arb (momentum, correlated pairs, regression analysis, news sentiment, etc). But now even the big stat arb strategies are running dry so people go further, chasing mirages (nonlinear regression, causality inference in large data sets, etc).
In modeling the market, it’s best to start with as much structure as possible before moving on to more amorphous statistical strategies. If you have to use statistical machine learning, encode as much trading domain knowledge as possible with specific distance/neighborhood metrics, linearity, variable importance weightings, hierarchy, low-dimensional factors, etc.
It’s good to have a heuristic feel for the …


I used to be very gung-ho about machine learning approaches to trading but I’m less so now. You have to understand that that there is a spectrum of alpha sources, from very specific structured arbitrage opportunities -> to stat arb -> to just voodoo nonsense.

As history goes on, hedge funds and other large players are absorbing the alpha from left to right. Having squeezed the pure arbs (ADR vs underlying, ETF vs components, mergers, currency triangles, etc) they then became hungry again and moved to stat arb (momentum, correlated pairs, regression analysis, news sentiment, etc). But now even the big stat arb strategies are running dry so people go further, chasing mirages (nonlinear regression, causality inference in large data sets, etc).
In modeling the market, it’s best to start with as much structure as possible before moving on to more amorphous statistical strategies. If you have to use statistical machine learning, encode as much trading domain knowledge as possible with specific distance/neighborhood metrics, linearity, variable importance weightings, hierarchy, low-dimensional factors, etc.
It’s good to have a heuristic feel for the danger/flexibility/noise sensitivity (synonyms) of each statistical learning tool. I roughly have this spectrum in my head:
Very specific, structured, safe
Optimize 1 parameter, require crossvalidation
↓
Optimize 2 parameters, require crossvalidation
↓
Optimize parameters with too little data, require regularization
↓
Extrapolation
↓
Nonlinear (SVM, tree bagging, etc)
↓
Higher-order variable dependencies
↓
Variable selection
↓
Structure learning
Very general, dangerous in noise, voodoo
This diagram is worth expanding. If anyone has any suggestions, please leave them.

More Read

Image
The Dirty (Not so Secret) Secret of IT Budgets
Write on The Emerging Role of the Analyst – SDC’s Analytics Blogarama Oct 6
Affiliate Summit West
Tactical Meandering
Metadata versus Taxonomy
TAGGED:data setsmachine learningmodeling
Share This Article
Facebook Pinterest LinkedIn
Share

Follow us on Facebook

Latest News

macro intelligence and ai
How Permutable AI is Advancing Macro Intelligence for Complex Global Markets
Artificial Intelligence Exclusive
warehouse accidents
Data Analytics and the Future of Warehouse Safety
Analytics Commentary Exclusive
stock investing and data analytics
How Data Analytics Supports Smarter Stock Trading Strategies
Analytics Exclusive
qr codes for data-driven marketing
Role of QR Codes in Data-Driven Marketing
Big Data Exclusive

Stay Connected

1.2KFollowersLike
33.7KFollowersFollow
222FollowersPin

You Might also Like

analyzing big data for its quality and value
Big Data

Use this Strategic Approach to Maximize Your Data’s Value

6 Min Read
AI is changing our lives in many ways
Artificial Intelligence

Artificial Intelligence Is Influencing Everyday Lives for the Better

5 Min Read
5 Ways Companies Use Machine Learning to Improve Workplace Productivity
Machine Learning

5 Ways Companies Use Machine Learning to Improve Workplace Productivity

6 Min Read
task management software
ExclusiveMachine LearningSoftware

A Guide To Machine Learning Foundations Of Task Management Software

6 Min Read

SmartData Collective is one of the largest & trusted community covering technical content about Big Data, BI, Cloud, Analytics, Artificial Intelligence, IoT & more.

data-driven web design
5 Great Tips for Using Data Analytics for Website UX
Big Data
ai is improving the safety of cars
From Bolts to Bots: How AI Is Fortifying the Automotive Industry
Artificial Intelligence

Quick Link

  • About
  • Contact
  • Privacy
Follow US
© 2008-25 SmartData Collective. All Rights Reserved.
Go to mobile version
Welcome Back!

Sign in to your account

Username or Email Address
Password

Lost your password?