Surgical or endovascular management of ruptured intracranial aneurysms: an agreement study
Bibliographic record
Abstract
OBJECTIVE: Ruptured intracranial aneurysms (RIAs) can be managed surgically or endovascularly. In this study, the authors aimed to measure the interobserver agreement in selecting the best management option for various patients with an RIA. METHODS: The authors constructed an electronic portfolio of 42 cases of RIA in which an angiographic image along with a brief clinical vignette for each patient were displayed. Undisclosed to the responders was that the RIAs had been categorized as International Subarachnoid Aneurysm Trial (ISAT) (small, anterior-circulation, non-middle cerebral artery location, n = 18) and non-ISAT (n = 22) aneurysms; the non-ISAT group also included 2 basilar apex aneurysms for which a high number of endovascular choices was expected. The portfolio was sent to 132 clinicians who manage patients with RIAs and circulated to members of an American surgical association. Judges were asked to choose between surgical and endovascular management, to indicate their level of confidence in the choice of treatment on a quantitative 0-10 scale, and to determine whether they would include the patient in a randomized trial in which both treatments are compared. Eleven clinicians were asked to respond twice at least 1 month apart. Responses were analyzed using kappa statistics. RESULTS: Eighty-five clinicians (58 cerebrovascular surgeons, 21 interventional neuroradiologists, and 6 interventional neurologists) answered the questionnaire. Overall, endovascular management was chosen more frequently (n = 2136 [59.8%] of 3570 answers). The proportions of decisions to clip were significantly higher for non-ISAT (50.8%) than for ISAT (26.2%) aneurysms (p = 0.0003). Interjudge agreement was only fair (kappa 0.210, 95% CI 0.158-0.276) for all cases and judges, despite high confidence levels (mean score > 8 for all cases). Agreement was no better within subgroups of clinicians with the same specialty, years of experience, or location of practice or across capability groups (ability to clip or coil, or both). When agreement was defined as > 80% of responders choosing the same option, agreement occurred for only 7 of 40 cases, all of which were ISAT aneurysms, for which coiling was preferred. CONCLUSIONS: Agreement between clinicians regarding the best management option was infrequent but centered around coiling for some ISAT aneurysms. Surgical clipping was chosen more frequently for non-ISAT aneurysms than for ISAT aneurysms. Patients with such an aneurysm might be candidates for inclusion in randomized trials.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".