Perceptions on Oral Ulcers From Facebook Page Categories: Observational Study
Bibliographic record
Abstract
BACKGROUND: Oral ulcers are a common condition affecting a considerable proportion of the population, and they are often associated with trauma and stress. They are very painful, and interfere with eating. As they are usually considered an annoyance, people may turn to social media for potential management options. Facebook is one of the most commonly accessed social media platforms and is the primary source of news information, including health information, for a significant percentage of American adults. Given the increasing importance of social media as a source of health information, potential remedies, and prevention strategies, it is essential to understand the type and quality of information available on Facebook regarding oral ulcers. OBJECTIVE: The goal of our study was to evaluate information on recurrent oral ulcers that can be accessed via the most popular social media network-Facebook. METHODS: We performed a keyword search of Facebook pages on 2 consecutive days in March 2022, using duplicate, newly created accounts, and then anonymized all posts. The collected pages were filtered, using predefined criteria to include only English-language pages wherein oral ulcer information was posted by the general public and to exclude pages created by professional dentists, associated professionals, organizations, and academic researchers. The selected pages were then screened for page origin and Facebook categories. RESULTS: Our initial keyword search yielded 517 pages; interestingly however, only 112 (22%) of pages had information relevant to oral ulcers, and 405 (78%) had irrelevant information, with ulcers being mentioned in relation to other parts of the human body. Excluding professional pages and pages without relevant posts resulted in 30 pages, of which 9 (30%) were categorized as "health/beauty" pages or as "product/service" pages, 3 (10%) were categorized as "medical & health" pages, and 5 (17%) were categorized as "community" pages. Majority of the pages (22/30, 73%) originated from 6 countries; most originated from the United States (7 pages), followed by India (6 pages). There was little information on oral ulcer prevention, long-term treatment, and complications. CONCLUSIONS: Facebook, in oral ulcer information dissemination, appears to be primarily used as an adjunct to business enterprises for marketing or for enhancing access to a product. Consequently, it was unsurprising that there was little information on oral ulcer prevention, long-term treatment, and complications. Although we made efforts to identify and select Facebook pages related to oral ulcers, we did not manually verify the authenticity or accuracy of the pages included in our analysis, potentially limiting the reliability of our findings or resulting in bias toward specific products or services. Although this work forms something of a pilot project, we plan to expand the project to encompass text mining for content analysis and include multiple social media platforms in the future.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.010 |
| Meta-epidemiology (narrow) | 0.000 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.002 | 0.002 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.002 | 0.002 |
| Open science | 0.000 | 0.002 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".