Text as data for evaluation: Natural language processing and large language models to generate novel insights from unstructured text data

3Citations
Citations of this article
19Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

Policy formulation and implementation generate large volumes of text. However, since reading all relevant sources is often impossible, evaluators must navigate the complexities of selecting the appropriate technology to efficiently extract meaningful information from growing amounts of unstructured text. Text mining blends interpretative and statistical methods to generate novel insights, potentially contributing to evidence-based policy-making. At the same time, biases, a potential lack of accuracy, explainability, and transparency create ethical concerns and make it necessary to combine natural language processing and human judgment to avoid over-reliance on the capabilities of these methods and, in particular, large language models. This article provides practical guidance on how evaluators can use natural language processing to convert unstructured data from text to structured data. It presents a decision framework that accounts for the characteristics of the data, the nature of the task, and the expected results, facilitating the selection of the appropriate technique.

Cite

CITATION STYLE

APA

Wencker, T., Borst-Graetz, J., & Niekler, A. (2025). Text as data for evaluation: Natural language processing and large language models to generate novel insights from unstructured text data. Evaluation, 31(3 Special Issue: Digitalization in Evaluations and Evaluations of Digitalization), 369–393. https://doi.org/10.1177/13563890251330911

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free