July 2020
Beginner to intermediate
820 pages
25h 30m
English
This is the first of three chapters dedicated to extracting signals for algorithmic trading strategies from text data using natural language processing (NLP) and machine learning (ML).
Text data is very rich in content but highly unstructured, so it requires more preprocessing to enable an ML algorithm to extract relevant information. A key challenge consists of converting text into a numerical format without losing its meaning. We will cover several techniques capable of capturing the nuances of language so that they can be used as input for ML algorithms.
In this chapter, we will introduce fundamental feature extraction techniques that focus on individual semantic units, that is, words or short ...
Read now
Unlock full access