MétaCan
Menu
Retour à la cohorte
Enregistrement W2899795704 · doi:10.2196/11232

The Future of Health Care: Protocol for Measuring the Potential of Task Automation Grounded in the National Health Service Primary Care System

2018· article· en· W2899795704 sur OpenAlexvenueno aff
Matthew Willis, Paul Duckworth, Angela Coulter, Eric T. Meyer, Michael A. Osborne

Notice bibliographique

RevueJMIR Research Protocols · 2018
Typearticle
Langueen
DomaineHealth Professions
ThématiqueElectronic Health Records Systems
Établissements canadiensnon disponible
Organismes subventionnairesnon disponible
Mots-clésAutomationTask (project management)Protocol (science)Work (physics)Health careService (business)Domain (mathematical analysis)Computer scienceKnowledge managementMedicineProcess managementData scienceEngineering managementEngineeringBusinessAlternative medicineSystems engineeringPolitical scienceMarketing

Résumé

récupéré en direct d'OpenAlex

BACKGROUND: Recent advances in technology have reopened an old debate on which sectors will be most affected by automation. This debate is ill served by the current lack of detailed data on the exact capabilities of new machines and how they are influencing work. Although recent debates about the future of jobs have focused on whether they are at risk of automation, our research focuses on a more fine-grained and transparent method to model task automation and specifically focus on the domain of primary health care. OBJECTIVE: This protocol describes a new wave of intelligent automation, focusing on the specific pressures faced by primary care within the National Health Service (NHS) in England. These pressures include staff shortages, increased service demand, and reduced budgets. A critical part of the problem we propose to address is a formal framework for measuring automation, which is lacking in the literature. The health care domain offers a further challenge in measuring automation because of a general lack of detailed, health care-specific occupation and task observational data to provide good insights on this misunderstood topic. METHODS: This project utilizes a multimethod research design comprising two phases: a qualitative observational phase and a quantitative data analysis phase; each phase addresses one of the two project aims. Our first aim is to address the lack of task data by collecting high-quality, detailed task-specific data from UK primary health care practices. This phase employs ethnography, observation, interviews, document collection, and focus groups. The second aim is to propose a formal machine learning approach for probabilistic inference of task- and occupation-level automation to gain valuable insights. Sensitivity analysis is then used to present the occupational attributes that increase/decrease automatability most, which is vital for establishing effective training and staffing policy. RESULTS: Our detailed fieldwork includes observing and documenting 16 unique occupations and performing over 130 tasks across six primary care centers. Preliminary results on the current state of automation and the potential for further automation in primary care are discussed. Our initial findings are that tasks are often shared amongst staff and can include convoluted workflows that often vary between practices. The single most used technology in primary health care is the desktop computer. In addition, we have conducted a large-scale survey of over 156 machine learning and robotics experts to assess what tasks are susceptible to automation, given the state-of-the-art technology available today. Further results and detailed analysis will be published toward the end of the project in early 2019. CONCLUSIONS: We believe our analysis will identify many tasks currently performed manually within primary care that can be automated using currently available technology. Given the proper implementation of such automating technologies, we expect considerable staff resources to be saved, alleviating some pressures on the NHS primary care staff. INTERNATIONAL REGISTERED REPORT IDENTIFIER (IRRID): DERR1-10.2196/11232.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,104
score de la tête « metaresearch » (Gemma)0,113
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Sans objet · Signal consensuel: Sans objet
GenreSignal candidat: Protocole · Signal consensuel: Protocole
Score de désaccord entre enseignants0,104
Score d'incertitude au seuil0,548

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,1040,113
Méta-épidémiologie (sens strict)0,0020,003
Méta-épidémiologie (sens large)0,0020,003
Bibliométrie0,0050,007
Études des sciences et des technologies0,0080,005
Communication savante0,0050,004
Science ouverte0,0050,006
Intégrité de la recherche0,0060,006
Charge utile insuffisante (le modèle a refusé de juger)0,0360,009

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,261
Tête enseignante GPT0,605
Écart entre enseignants0,343 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeSans objet
Domainenon disponible
GenreProtocole

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations19
Publié2018
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueJMIR Research ProtocolsMême sujetElectronic Health Records SystemsTravaux en français237 207