Revising <i>Diagnostic and Statistical Manual of Mental Disorders</i> , Fifth Edition, criteria for the bipolar disorders: Phase I of the AREDOC project
Bibliographic record
Abstract
OBJECTIVE: To derive new criteria sets for defining manic and hypomanic episodes (and thus for defining the bipolar I and II disorders), an international Task Force was assembled and termed AREDOC reflecting its role of Assessment, Revision and Evaluation of DSM and other Operational Criteria. This paper reports on the first phase of its deliberations and interim criteria recommendations. METHOD: , and recent International Classification of Diseases criteria, identifying their limitations and generating modified criteria sets for further in-depth consideration. Task Force members responded to recommendations for modifying criteria and from these the most problematic issues were identified. RESULTS: Principal issues focussed on by Task Force members were how best to differentiate mania and hypomania, how to judge 'impairment' (both in and of itself and allowing that functioning may sometimes improve during hypomanic episodes) and concern that rejecting some criteria (e.g. an imposed duration period) might risk false-positive diagnoses of the bipolar disorders. CONCLUSION: This first-stage report summarises the clinical opinions of international experts in the diagnosis and management of the bipolar disorders, allowing readers to contemplate diagnostic parameters that may influence their clinical decisions. The findings meaningfully inform subsequent Task Force stages (involving a further commentary stage followed by an empirical study) that are expected to generate improved symptom criteria for diagnosing the bipolar I and II disorders with greater precision and to clarify whether they differ dimensionally or categorically.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".