Linguistic Analysis of Online Communication About a Novel Persecutory Belief System (Gangstalking): Mixed Methods Study
Bibliographic record
Abstract
BACKGROUND: Gangstalking is a novel persecutory belief system whereby those affected believe they are being followed, stalked, and harassed by a large number of people, often numbering in the thousands. The harassment is experienced as an accretion of innumerable individually benign acts such as people clearing their throat, muttering under their breath, or giving dirty looks as they pass on the street. Individuals affected by this belief system congregate in online fora to seek support, share experiences, and interact with other like-minded individuals. Such people identify themselves as targeted individuals. OBJECTIVE: The objective of the study was to characterize the linguistic and rhetorical practices used by contributors to the gangstalking forum to construct, develop, and contest the gangstalking belief system. METHODS: This mixed methods study employed corpus linguistics, which involves using computational techniques to examine recurring linguistic patterns in large, digitized bodies of authentic language data. Discourse analysis is an approach to text analysis which focuses on the ways in which linguistic choices made by text creators contribute to particular functions and representations. We assembled a 225,000-word corpus of postings on a gangstalking support forum. We analyzed these data using keyword analysis, collocation analysis, and manual examination of concordances to identify discursive and rhetorical practices among self-identified targeted individuals. RESULTS: The gangstalking forum served as a site of discursive contest between 2 opposing worldviews. One is that gangstalking is a widespread, insidious, and centrally coordinated system of persecution employing community members, figures of authority, and state actors. This was the dominant discourse in the study corpus. The opposing view is a medicalized discourse supporting gangstalking as a form of mental disorder. Contributors used linguistic practices such as presupposition, nominalization, and the use of specialized jargon to construct gangstalking as real and external to the individual affected. Although contributors generally rejected the notion that they were affected by mental disorder, in some instances, they did label others in the forum as impacted/affected by mental illness if their accounts if their accounts were deemed to be too extreme or bizarre. Those affected demonstrated a concern with accumulating evidence to prove their position to incredulous others. CONCLUSIONS: The study found that contributors to the study corpus accomplished a number of tasks. They used linguistic practices to co-construct an internally coherent and systematized persecutory belief system. They advanced a position that gangstalking is real and contested the medicalizing discourse that gangstalking is a form of mental disorder. They supported one another by sharing similar experiences and providing encouragement and advice. Finally, they commiserated over the challenges of proving the existence of gangstalking.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.031 | 0.014 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".