Abstract
This paper takes a corpus-based approach to study vulgar language in online communication across 20 English-speaking regions based on the Global Web-Based English Corpus (GloWbE). The identification of vulgarity combines word lists used in profanity detection with regular expressions to identify a wide range of vulgar elements including spelling variants and obscured forms. The results show a notable trend for inner circle L1-varieties to exhibit higher rates of vulgarity online compared to outer circle and L2-varieties. The results also show that inner circle varieties have lower adapted corrected type-token rations which indicates that inner circle variety speakers use more varied English vulgar forms compared with speakers from other circle varieties. In addition, there is a general register difference with vulgarity being more common in blog data compared with general web content. Finally, the results show that different regions exhibit preferences for specific vulgar lemmas feck being preferred in Ireland, cunt, in Britain, and ass(hole) in the United States. The findings are interpreted to show that cultural differences are reflected in region-specific preferences for vulgarity and that the creativity observed in inner circle varieties is linked to norm-setting compared to norm-reception associated with outer circle varieties.
Author supplied keywords
Cite
CITATION STYLE
Schweinberger, M., & Burridge, K. (2025). Vulgarity in online discourse around the English-speaking world. Lingua, 321. https://doi.org/10.1016/j.lingua.2025.103946
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.