An organization‐ and category‐level comparison of diagnostic requirements for mental disorders in <scp>ICD</scp>‐11 and <scp>DSM</scp>‐5
Bibliographic record
Abstract
In 2013, the American Psychiatric Association (APA) published the 5th edition of its Diagnostic and Statistical Manual of Mental Disorders (DSM-5). In 2019, the World Health Assembly approved the 11th revision of the International Classification of Diseases (ICD-11). It has often been suggested that the field would benefit from a single, unified classification of mental disorders, although the priorities and constituencies of the two sponsoring organizations are quite different. During the development of the ICD-11 and DSM-5, the World Health Organization (WHO) and the APA made efforts toward harmonizing the two systems, including the appointment of an ICD-DSM Harmonization Group. This paper evaluates the success of these harmonization efforts and provides a guide for practitioners, researchers and policy makers describing the differences between the two systems at both the organizational and the disorder level. The organization of the two classifications of mental disorders is substantially similar. There are nineteen ICD-11 disorder categories that do not appear in DSM-5, and seven DSM-5 disorder categories that do not appear in the ICD-11. We compared the Essential Features section of the ICD-11 Clinical Descriptions and Diagnostic Guidelines (CDDG) with the DSM-5 criteria sets for 103 diagnostic entities that appear in both systems. We rated 20 disorders (19.4%) as having major differences, 42 disorders (40.8%) as having minor definitional differences, 10 disorders (9.7%) as having minor differences due to greater degree of specification in DSM-5, and 31 disorders (30.1%) as essentially identical. Detailed descriptions of the major differences and some of the most important minor differences, with their rationale and related evidence, are provided. The ICD and DSM are now closer than at any time since the ICD-8 and DSM-II. Differences are largely based on the differing priorities and uses of the two diagnostic systems and on differing interpretations of the evidence. Substantively divergent approaches allow for empirical comparisons of validity and utility and can contribute to advances in the field.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.010 | 0.037 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.006 | 0.006 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.002 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".