Visualizing YouTube Commenters’ Conceptions of the US Health Care System: Semantic Network Analysis Method for Evidence-Based Policy Making
Bibliographic record
Abstract
BACKGROUND: The challenge of extracting meaningful patterns from the overwhelming noise of social media to guide decision-makers remains largely unresolved. OBJECTIVE: This study aimed to evaluate the application of a semantic network method for creating an interactive visualization of social media discourse surrounding the US health care system. METHODS: Building upon bibliometric approaches to conducting health studies, we repurposed the VOSviewer software program to analyze 179,193 YouTube comments about the US health care system. Using the overlay-enhanced semantic network method, we mapped the contents and structure of the commentary evoked by 53 YouTube videos uploaded in 2014 to 2023 by right-wing, left-wing, and centrist media outlets. The videos included newscasts, full-length documentaries, political satire, and stand-up comedy. We analyzed term co-occurrence network clusters, contextualized with custom-built information layers called overlays, and performed tests of the semantic network's robustness, representativeness, structural relevance, semantic accuracy, and usefulness for decision support. We examined how the comments mentioning 4 health system design concepts-universal health care, Medicare for All, single payer, and socialized medicine-were distributed across the network terms. RESULTS: Grounded in the textual data, the macrolevel network representation unveiled complex discussions about illness and wellness; health services; ideology and society; the politics of health care agendas and reforms, market regulation, and health insurance; the health care workforce; dental care; and wait times. We observed thematic alignment between the network terms, extracted from YouTube comments, and the videos that elicited these comments. Discussions about illness and wellness persisted across time, as well as international comparisons of costs of ambulances, specialist care, prescriptions, and appointment wait times. The international comparisons were linked to commentaries with a higher concentration of British-spelled words, underscoring the global nature of the US health care discussion, which attracted domestic and global YouTube commenters. Shortages of nurses, nurse burnout, and their contributing factors (eg, shift work, nurse-to-patient staffing ratios, and corporate greed) were covered in comments with many likes. Comments about universal health care had much higher use of ideological terms than comments about single-payer health systems. CONCLUSIONS: YouTube users addressed issues of societal and policy relevance: social determinants of health, concerns for populations considered vulnerable, health equity, racism, health care quality, and access to essential health services. Versatile and applicable to health policy studies, the method presented and evaluated in our study supports evidence-based decision-making and contextualized understanding of diverse viewpoints. Interactive visualizations can help to uncover large-scale patterns and guide strategic use of analytical resources to perform qualitative research.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.003 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.002 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".