Abstract
This article investigates the authorship question surrounding William Shakespeare’s works using a novel approach called the ‘Deep Impostor’ methodology. The approach uses a set of known impostor texts to analyze the origin of a target text collection. Both the target texts and impostors are divided into an equal number of word segments. A deep neural network, either a Convolutional Neural Network (CNN) or a pre-trained BERT transformer, is then trained and fine-tuned to differentiate between impostor segments. Once assigned, each target text is transformed into a numerical signal by averaging its segment assignments. The Dynamic Time Warping distance is afterward evaluated between these signals to measure their similarity. The Isolation Forest algorithm identifies outliers within the target text collection for each impostor pair by assigning appropriate scores to each tested text. In the summarizing step, the tested creations are clustered into two groups. In the case of a CNNbased model, the first resulting cluster contains fifteen creations. General evaluations lead to the conclusion to suggest that these are not authored by Shakespeare. The remaining documents are classified as authentic Shakespearean works.
Author supplied keywords
Cite
CITATION STYLE
Volkovich, Z., & Avros, R. (2025). Comprehension of the Shakespeare authorship question through deep impostors approach. Digital Scholarship in the Humanities, 40(1), 308–328. https://doi.org/10.1093/llc/fqaf009
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.