MAEC: A Multimodal Aligned Earnings Conference Call Dataset for Financial Risk Prediction

50Citations
Citations of this article
45Readers
Mendeley users who have this article in their library.
Get full text

Abstract

In the area of natural language processing, various financial datasets have informed recent research and analysis including financial news, financial reports, social media, and audio data from earnings calls. We introduce a new, large-scale multi-modal, text-audio paired, earnings-call dataset named MAEC, based on S&P 1500 companies. We describe the main features of MAEC, how it was collected and assembled, paying particular attention to the text-audio alignment process used. We present the approach used in this work as providing a suitable framework for processing similar forms of data in the future. The resulting dataset is more than six times larger than those currently available to the research community and we discuss its potential in terms of current and future research challenges and opportunities. All resources of this work are available at https://github.com/Earnings-Call-Dataset/

Cite

CITATION STYLE

APA

Li, J., Yang, L., Smyth, B., & Dong, R. (2020). MAEC: A Multimodal Aligned Earnings Conference Call Dataset for Financial Risk Prediction. In International Conference on Information and Knowledge Management, Proceedings (pp. 3063–3070). Association for Computing Machinery. https://doi.org/10.1145/3340531.3412879

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free