Automated Generation and Evaluation of Interactive-Fiction Serious Games with Open-Weight LLMs

0Citations
Citations of this article
N/AReaders
Mendeley users who have this article in their library.

Abstract

This work investigates whether open-weight large language models can automatically generate runnable and educationally faithful serious games in a constrained, text-only interactive-fiction (IF) setting. The target games are station-based single-player serious games for knowledge assessment, implemented as IF in a structured, machine-readable text format, and used here as a first step towards later ambient scenarios. A fully automated pipeline called SINE (Serious Interactive Narrative Engine) is evaluated with four prompting strategies, grammar-guided decoding, deterministic validation, and a repair agent. Across a staged evaluation with 240 seeds and increasing complexity, finalist configurations reach success rates between roughly 68% and 86% on the joint criterion of compilation, playability, and learning-goal fidelity. Repair iterations proved central to robustness, whereas grammar masking on top of reasoning prompts did not consistently improve outcomes. The study provides a reproducible benchmark setup, open artifacts, and a constrained generation pipeline as a basis for later extensions toward broader serious game scenarios.

Cite

CITATION STYLE

APA

Rogosch, F., & Schrader, A. (2026). Automated Generation and Evaluation of Interactive-Fiction Serious Games with Open-Weight LLMs. Applied Sciences (Switzerland), 16(6). https://doi.org/10.3390/app16062932

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free