Topic modeling has been widely adopted by researchers for a variety of different research problems that involve the mining of text corpora to generate a latent set of topics. Specifically, the Latent Dirichlet Allocation (LDA) algorithm is well documented within academic literature in terms of its application and automated topic generation from data sources such as blogs, social media, and other text collections. YouTube now offers access to over a billion auto-generated video transcript documents that have been recorded and posted to its social platform. The availability of this data offers an opportunity for researchers to investigate a variety of topics that are being discussed and posted to the platform. Specifically, we will study, using the LDA algorithm, discussions related to emerging technologies that have been posted on YouTube to better understand what latent topics can be auto-generated and what kind of methodology can be used to analyze this data.
CITATION STYLE
Daniel, C., & Dutta, K. (2018). Automated generation of latent topics on emerging technologies from YouTube video content. In Proceedings of the Annual Hawaii International Conference on System Sciences (Vol. 2018-January, pp. 1762–1770). IEEE Computer Society. https://doi.org/10.24251/hicss.2018.222
Mendeley helps you to discover research relevant for your work.