MétaCan
Menu
← Back to cohort
Record W4411383563 · doi:10.2196/72115

Data Mining–Based Model for Computer-Aided Diagnosis of Autism and Gelotophobia: Mixed Methods Deep Learning Approach

2025· article· en· W4411383563 on OpenAlexvenueno aff
Mohamed Eldawansy, Hazem M. El‐Bakry, Samaa M. Shohieb

Bibliographic record

VenueJMIR Formative Research · 2025
Typearticle
Languageen
FieldNeuroscience
TopicAutism Spectrum Disorder Research
Canadian institutionsnot available
Fundersnot available
KeywordsPreprintArtificial intelligenceComputer sciencePsychologyMachine learningData scienceWorld Wide Web

Abstract

fetched live from OpenAlex

BACKGROUND: Gelotophobia, the fear of being laughed at, is a social anxiety condition that affects approximately 6% of neurotypical individuals and up to 45% of those with autism spectrum disorder (ASD). This comorbidity can significantly impair the quality of life, particularly in adolescents with high-functioning ASD, where the prevalence reaches 41.98%. Accurate and automated detection tools could enhance early diagnosis and intervention. OBJECTIVE: This study aimed to develop a deep learning-based diagnostic system that integrates facial emotion recognition with validated questionnaires to detect gelotophobia in individuals with or without ASD. METHODS: The system was trained to identify ASD status using a balanced dataset of 2932 facial images (n=1466; 50% from individuals with ASD and n=1466; 50% from neurotypical individuals). The images were processed using the DeepFace library to extract facial features, which were then used as input for the deep learning classifier. After identifying ASD status, the same images were further analyzed using the pretrained DeepFace model to evaluate facial expressions for signs of gelotophobia. In cases where facial cues were ambiguous, the GELOPH<15> questionnaire, consisting of 15 items, was administered to confirm the diagnosis The system was fully implemented using the Python programming language. Deep learning models were developed using libraries such as PyTorch for training the multilayer perceptron classifier, while CUDA was used to accelerate computations on compatible graphics processing units. Additional libraries from the Python programming language, such as scikit-learn, NumPy, and Pandas, were used for preprocessing, model evaluation, and data manipulation. DeepFace was integrated using its Python application programming interface for facial recognition and emotion classification. RESULTS: The dataset comprised 2932 facial images collected from platforms such as Kaggle and ASD-related websites, including 1466 (50%) images of children with ASD and 1466 (50%) images of neurotypical children. The dataset was split into 2653 (90.48%) training samples and 279 (9.51%) testing samples, with each image contributing 100,352 extracted features. We applied various machine learning models for ASD identification. The system achieved an overall prediction accuracy of 92% across both training and testing datasets, with the multilayer perceptron model demonstrating the highest testing accuracy. The system successfully classified gelotophobia in cases where facial expressions were clear. However, in cases of ambiguous facial cues, the DeepFace model alone was insufficient. Incorporating the GELOPH<15> questionnaire improved diagnostic reliability and consistency. CONCLUSIONS: This study demonstrates the effectiveness of combining deep learning techniques with validated diagnostic tools for detecting gelotophobia, particularly in individuals with ASD. The high accuracy achieved highlights the system's potential for clinical and research applications, contributing to the improved understanding and management of gelotophobia among groups considered socially vulnerable. Future research could expand the system's applications to broader psychological assessments.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.002
metaresearch head score (Gemma)0.003
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Simulation or modeling · Consensus signal: Simulation or modeling
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.013
Threshold uncertainty score0.027

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0020.003
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0010.002
Bibliometrics0.0010.001
Science and technology studies0.0000.000
Scholarly communication0.0010.001
Open science0.0020.001
Research integrity0.0020.002
Insufficient payload (model declined to judge)0.0020.001

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.188
GPT teacher head0.479
Teacher spread0.291 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designSimulation or modeling
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations1
Published2025
Admission routes1
Has abstractyes

Explore more

Same venueJMIR Formative Research→Same topicAutism Spectrum Disorder Research→French-language works237,207→