We introduce a new text-mining methodology that extracts sentiment information from news articles to predict asset returns. Unlike more common sentiment scores used for stock return prediction (e.g., those sold by commercial vendors or built with dictionary-based methods), our supervised learning framework constructs a sentiment score that is specifically adapted to the problem of return prediction. Our method proceeds in three steps: 1) isolating a list of sentiment terms via predictive screening, 2) assigning sentiment weights to these words via topic modeling, and 3) aggregating terms into an article-level sentiment score via penalized likelihood. We derive theoretical guarantees on the accuracy of estimates from our model with minimal assumptions. In our empirical analysis, we text-mine one of the most actively monitored streams of news articles in the financial system—the Dow Jones Newswires—and show that our supervised sentiment model excels at extracting return-predictive signals in this context.

More on this topic

BFI Working Paper·Feb 10, 2025

Policy Interventions and China’s Stock Market in the Early Stages of the COVID-19 Pandemic

Steven Davis, Dingqian Liu, Xuguang Simon Sheng, and Yan Wang
Topics: COVID-19, Financial Markets
BFI Working Paper·Feb 3, 2025

The Long and Short of Financial Development

Douglas W. Diamond, Yunzhi Hu, and Raghuram Rajan
Topics: Financial Markets
BFI Working Paper·Jan 28, 2025

An Informationally-Robust Market Model of Perfect Competition

Benjamin Brooks, Songzi Du, and Linchen Zhang
Topics: Financial Markets