Data Mining–Based Model for Computer-Aided Diagnosis of Autism and Gelotophobia: Mixed Methods Deep Learning Approach
Notice bibliographique
Résumé
BACKGROUND: Gelotophobia, the fear of being laughed at, is a social anxiety condition that affects approximately 6% of neurotypical individuals and up to 45% of those with autism spectrum disorder (ASD). This comorbidity can significantly impair the quality of life, particularly in adolescents with high-functioning ASD, where the prevalence reaches 41.98%. Accurate and automated detection tools could enhance early diagnosis and intervention. OBJECTIVE: This study aimed to develop a deep learning-based diagnostic system that integrates facial emotion recognition with validated questionnaires to detect gelotophobia in individuals with or without ASD. METHODS: The system was trained to identify ASD status using a balanced dataset of 2932 facial images (n=1466; 50% from individuals with ASD and n=1466; 50% from neurotypical individuals). The images were processed using the DeepFace library to extract facial features, which were then used as input for the deep learning classifier. After identifying ASD status, the same images were further analyzed using the pretrained DeepFace model to evaluate facial expressions for signs of gelotophobia. In cases where facial cues were ambiguous, the GELOPH<15> questionnaire, consisting of 15 items, was administered to confirm the diagnosis The system was fully implemented using the Python programming language. Deep learning models were developed using libraries such as PyTorch for training the multilayer perceptron classifier, while CUDA was used to accelerate computations on compatible graphics processing units. Additional libraries from the Python programming language, such as scikit-learn, NumPy, and Pandas, were used for preprocessing, model evaluation, and data manipulation. DeepFace was integrated using its Python application programming interface for facial recognition and emotion classification. RESULTS: The dataset comprised 2932 facial images collected from platforms such as Kaggle and ASD-related websites, including 1466 (50%) images of children with ASD and 1466 (50%) images of neurotypical children. The dataset was split into 2653 (90.48%) training samples and 279 (9.51%) testing samples, with each image contributing 100,352 extracted features. We applied various machine learning models for ASD identification. The system achieved an overall prediction accuracy of 92% across both training and testing datasets, with the multilayer perceptron model demonstrating the highest testing accuracy. The system successfully classified gelotophobia in cases where facial expressions were clear. However, in cases of ambiguous facial cues, the DeepFace model alone was insufficient. Incorporating the GELOPH<15> questionnaire improved diagnostic reliability and consistency. CONCLUSIONS: This study demonstrates the effectiveness of combining deep learning techniques with validated diagnostic tools for detecting gelotophobia, particularly in individuals with ASD. The high accuracy achieved highlights the system's potential for clinical and research applications, contributing to the improved understanding and management of gelotophobia among groups considered socially vulnerable. Future research could expand the system's applications to broader psychological assessments.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,002 | 0,003 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,001 |
| Méta-épidémiologie (sens large) | 0,001 | 0,002 |
| Bibliométrie | 0,001 | 0,001 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,001 | 0,001 |
| Science ouverte | 0,002 | 0,001 |
| Intégrité de la recherche | 0,002 | 0,002 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,002 | 0,001 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».