MétaCan
Menu
Back to cohort
Record W4404738142 · doi:10.1016/j.anplas.2024.11.002

La critique d’un outil d’intelligence artificielle dans l’évaluation des paralysies faciales périphériques

2024· article· fr· W4404738142 on OpenAlexafffund
Hélène Kerleau, Lionel Perrin, Karine Marcotte, Sarah Martineau

Bibliographic record

VenueAnnales de Chirurgie Plastique Esthétique · 2024
Typearticle
Languagefr
FieldMedicine
TopicFacial Nerve Paralysis Treatment and Research
Canadian institutionsHôpital Maisonneuve-RosemontCentre Intégré Universitaire de Santé et de Services Sociaux du Centre-Sud-de-l'Île-de-MontréalHôpital du Sacré-Cœur de Montréal
FundersFonds de Recherche du Québec - Santé
KeywordsMedicineHumanitiesPhilosophy

Abstract

fetched live from OpenAlex

La paralysie faciale périphérique (PFP) est une altération du fonctionnement de certains muscles du visage, suite à une lésion du nerf facial. Cette pathologie entraîne des conséquences fonctionnelles et esthétiques qui impactent la qualité de vie des individus. Leur prise en soin est donc cruciale et débute par une évaluation précise. Actuellement, on utilise principalement des échelles de notation telles que Sunnybrook Facial Grading System (SFGS) ou House-Brackmann Grading System (HBGS), basées sur le jugement du clinicien. Cependant, ces méthodes d’évaluation laissent place à une certaine subjectivité. Grâce aux récentes avancées technologiques, on s’intéresse davantage à l’intelligence artificielle (IA). L’IA pourrait permettre de développer un outil d’évaluation objectif, automatisé et quantitatif, applicable en milieu clinique. Cette approche vise à diminuer la subjectivité induite par les évaluations actuelles. Nous avons mené une étude rétrospective auprès de 38 patients présentant une PFP modérée-sévère à totale. L’objectif de l’étude était de recenser les bénéfices et les limites d’Emotrics+, un logiciel de métriques faciales basé sur l’IA, afin de déterminer si cet outil est applicable en clinique. Ce protocole s’est déroulé à deux périodes différentes (14 jours et 1 an post-PFP) en utilisant l’échelle SFGS et le logiciel Emotrics+. Nous avons évalué les fiabilités inter-juges et intra-juge afin de déterminer la fiabilité et la reproductibilité des deux outils. Puis, nous avons établi une corrélation entre les deux outils pour déterminer si Emotrics+ suivait la tendance de SFGS. Nos résultats actuels ne justifient pas l’applicabilité immédiate de cet outil. Cependant, avec des ajustements appropriés, Emotrics+ présente un potentiel certain. Peripheral facial palsy (PFP) is an alteration in the functioning of some facial muscles following an injury to the facial nerve. This pathology has functional and aesthetic consequences that impact the quality of life of patients. Their care is essential and begins with an accurate assessment. Currently, scoring scales such as Sunnybrook Facial Grading System (SFGS) or House-Brackmann Grading System (HBGS) are used, based on clinician judgment. However, these evaluation methods can be subject to a certain degree of subjectivity. Recent advances in technology have led to increased interest in artificial intelligence (AI). AI could make it possible to develop an objective, automated and quantitative assessment tool, applicable in a clinical setting. This approach aims to reduce the subjectivity induced by current evaluation. We conducted a retrospective study of 38 patients with moderate-severe to total PFPs. The objective of the study is to identify the benefits and limitations of Emotrics+, a facial metrics tool based on AI, in order to determine whether the tool is applicable in the clinic. This protocol took place at two different time periods (14 days and 1 year post-PFP) using the SFGS scale and the Emotrics+ software. We evaluated the inter-rater and intra-rater reliability in order to determine the reliability and the reproducibility of the two tools. Then, we established a correlation between the two tools to determine if Emotrics+ followed SFGS's trend. Our currents results do not support the immediate applicability of this software. However, with appropriates adjustments, Emotrics+ has a certain potential.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.003
metaresearch head score (Gemma)0.001
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMeta-epidemiology (narrow), Insufficient payload (model declined to judge)
Consensus categoriesInsufficient payload (model declined to judge)
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Other design · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.754
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0030.001
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0010.001
Bibliometrics0.0010.001
Science and technology studies0.0010.002
Scholarly communication0.0010.001
Open science0.0010.000
Research integrity0.0010.002
Insufficient payload (model declined to judge)0.0010.002

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.046
GPT teacher head0.357
Teacher spread0.310 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; both teacher heads agree on what is shown here.

Study designOther design
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations1
Published2024
Admission routes2
Has abstractyes

Explore more

Same venueAnnales de Chirurgie Plastique EsthétiqueSame topicFacial Nerve Paralysis Treatment and ResearchFrench-language works237,207