How Do Diabetes Models Measure Up? A Review of Diabetes Economic Models and ADA Guidelines
Bibliographic record
Abstract
Introduction: Introduction: Economic models and computer simulation models have been used for assessing short-term cost-effectiveness of interventions and modelling long-term outcomes and costs. Several guidelines and checklists have been published to improve the methods and reporting. This article presents an overview of published diabetes models with a focus on how well the models are described in relation to the considerations described by the American Diabetes Association (ADA) guidelines. Methods: Relevant electronic databases and National Institute for Health and Care Excellence (NICE) guidelines were searched in December 2012. Studies were included in the review if they estimated lifetime outcomes for patients with type 1 or type 2 diabetes. Only unique models, and only the original papers were included in the review. If additional information was reported in subsequent or paired articles, then additional citations were included. References and forward citations of relevant articles, including the previous systematic reviews were searched using a similar method to pearl growing. Four principal areas were included in the ADA guidance reporting for models: transparency, validation, uncertainty, and diabetes specific criteria. Results: A total of 19 models were included. Twelve models investigated type 2 diabetes, two developed type 1 models, two created separate models for type 1 and type 2, and three developed joint type 1 and type 2 models. Most models were developed in the United States, United Kingdom, Europe or Canada. Later models use data or methods from earlier models for development or validation. There are four main types of models: Markov-based cohort, Markov-based microsimulations, discrete-time microsimulations, and continuous time differential equations. All models were long-term diabetes models incorporating a wide range of compilations from various organ systems. In early diabetes modelling, before the ADA guidelines were published, most models did not include descriptions of all the diabetes specific components of the ADA guidelines but this improved significantly by 2004. Conclusion: A clear, descriptive short summary of the model was often lacking. Descriptions of model validation and uncertainty were the most poorly reported of the four main areas, but there exist conferences focussing specifically on the issue of validation. Interdependence between the complications was the least well incorporated or reported of the diabetes-specific criterion.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.039 | 0.142 |
| Meta-epidemiology (narrow) | 0.003 | 0.002 |
| Meta-epidemiology (broad) | 0.007 | 0.009 |
| Bibliometrics | 0.017 | 0.016 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.006 | 0.006 |
| Open science | 0.005 | 0.002 |
| Research integrity | 0.004 | 0.006 |
| Insufficient payload (model declined to judge) | 0.004 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".