MétaCan
Menu
Retour à la cohorte
Enregistrement W4414761919 · doi:10.1101/2025.09.30.25336981

British news media representations of mpox during the 2022 and 2024 outbreaks: a mixed-methods analysis using corpus linguistics

2025· preprint· en· W4414761919 sur OpenAlexaff
Beth Malory, Marthe Le Prevost, Emily Jay Nicholls, Davide Bilardi, Shema Tariq

Notice bibliographique

RevuemedRxiv · 2025
Typepreprint
Langueen
DomaineImmunology and Microbiology
ThématiquePoxvirus research and outbreaks
Établissements canadiensCentre for Global Health Research
Organismes subventionnairesEuropean Commission
Mots-clésCorpus linguisticsBlameNews mediaLinguistic analysisAttributionNews valuesDiscourse analysisHeadlinePublic discourse

Résumé

récupéré en direct d'OpenAlex

Abstract Background Since 2022, over 100,000 people across 100 countries have been diagnosed with mpox (formerly monkeypox, renamed by the World Health Organization (WHO) in November 2022). News media plays a central role in outbreaks, disseminating information and shaping public discourse. Corpus linguistic approaches to massive language datasets can reveal how such outbreaks are represented but remain under-used in public health. Using these methods, we investigated representations of mpox in British news media during the 2022 and 2024 outbreaks. Methods We analysed the 83-billion-word English Trends corpus in SketchEngine, quantifying use of “ monkeypox ” and “ mpox” in 2022–2024, and applying Corpus-Assisted Discourse Studies to compare news media content from 2022 and 2024 (peak incidence periods), comprising 1.2 billion and 500 million words, respectively. Using corpus linguistic tools, we explored the contexts in which mpox and monkeypox occurred, assessing shifts in representation. Findings Monthly use of “monkeypox” peaked at 0.07 occurrences per million words (n=6591) in May 2022, dropping by 99.8% by November 2022. Grammatical and lexical analysis of 2022 reporting found frequent attribution blame for transmission, particularly to gay and bisexual men who have sex with men (GBMSM). In 2024, coverage adopted more neutral language, largely avoiding stigmatisation. Interpretation UK news media reporting on mpox shifted from stigmatising language in 2022, often targeting GBMSM, to more neutral and inclusive coverage in 2024. The WHO-endorsed nomenclature change may have contributed, illustrating the impact of such interventions. This study demonstrates the value of corpus methods in tracking linguistic representations of infectious disease outbreaks. Funding This work is part of the VERDI project (101045989) which is funded by the European Union. Views and opinions expressed are however those of the authors only and do not necessarily reflect those of the European Union. BM, MLP and ST are partly funded through a Wellcome Accelerator Award held by ST (316319/Z/24/Z). Research in context Evidence before this study We reviewed existing literature on media representations of mpox, and other infectious diseases, focusing on stigma, framing, and public perception. We searched academic databases, including studies that examined media discourse and linguistic framing, and those using qualitative or corpus-based methods. While some explored stigma in mpox media coverage, few studies applied computational (computer-based) linguistic analysis to large-scale media datasets, or compared media narratives across different outbreak periods. Added value of this study This study is the first to use a ‘big data’ approach to explore representations of mpox. It uses corpus linguistic methods to analyse over two billion words of British news media content across two mpox outbreaks. It reveals a shift from stigmatising language in 2022—often targeting specific communities—to more neutral and inclusive reporting in 2024. The findings demonstrate how media language evolves in response to public health guidance and highlight the potential of corpus linguistic methods to uncover patterns in public discourse. Implications of all the available evidence Media language plays a powerful role in shaping public understanding and attitudes during health emergencies. This study shows that changes in terminology and framing can reduce stigma and improve public health communication. These insights support the need for proactive media guidance and the use of linguistic analysis to inform future policy, practice, and research in outbreak response and health communication.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,006
score de la tête « metaresearch » (Gemma)0,032
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Qualitatif · Signal consensuel: aucune
GenreSignal candidat: Empirique · Signal consensuel: Empirique
Score de désaccord entre enseignants0,114
Score d'incertitude au seuil0,227

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,0060,032
Méta-épidémiologie (sens strict)0,0000,000
Méta-épidémiologie (sens large)0,0000,001
Bibliométrie0,0070,010
Études des sciences et des technologies0,0020,001
Communication savante0,0030,001
Science ouverte0,0010,002
Intégrité de la recherche0,0010,001
Charge utile insuffisante (le modèle a refusé de juger)0,0040,001

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,024
Tête enseignante GPT0,348
Écart entre enseignants0,324 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeQualitatif
Domainenon disponible
GenreEmpirique

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations0
Publié2025
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revuemedRxivMême sujetPoxvirus research and outbreaksTravaux en français237 207