Indico Data releases industry-first large language model benchmark for document understanding tasks
Learn More
  Everest Group IDP
             PEAK Matrix® 2022  
Indico Named as Major Contender and Star Performer in Everest Group's PEAK Matrix® for Intelligent Document Processing (IDP)
Access the Report


Is feature selection needed to be performed in lexicon-based sentiment analysis?

March 27, 2018 | Ask Slater

Back to Blog

 The two are typically at odds. Generally speaking your lexicon defines your features. So you’re performing feature selection as you build the lexicon and once you have your lexicon you stop choosing new features.

If you’re starting with a lexicon and want to add features you could add them to the lexicon you’re working with, or attempt to create some kind of ensemble approach that combines the features in your lexicon with whatever features you create yourself.

However, for sentiment analysis you really shouldn’t be manually creating features, or using lexicons. These approaches generally lead to extremely brittle models with very poor performance. The problem is that if you are engineering your features or changing your lexicon in response to test errors then you’re manually overfitting.

View original question on Quora >

Follow Slater on Quora >>

Increase intake capacity. Drive top line revenue growth.


Get started with Indico

1-1 Demo



Gain insights from experts in automation, data, machine learning, and digital transformation.

Unstructured Unlocked

Enterprise leaders discuss how to unlock value from unstructured data.

YouTube Channel

Check out our YouTube channel to see clips from our podcast and more.
Subscribe to our blog

Get our best content on intelligent automation sent to your inbox weekly!