Intra-rater reliability of the bath ankylosing spondylitis disease activity index (BASDAI) and the bath ankylosing spondylitis functional index (BASFI) in children with spondyloarthritis
Bibliographic record
Abstract
Patients diagnosed with ERA (ILAR criteria) and followed at The Hospital for Sick Children Spondyloarthritis Clinic were included in this study. Patients were excluded if they were unable to understand, speak or write English, if they were less than 6 years of age or greater than 18 years of age. Prospective subjects were consecutively enrolled ( June 2009 to June 2010) and, the patient and/or one parent completed the BASDAI and BASFI at baseline and 2 weeks later (a period during which little change is expected). Intra-class correlation coefficient (ICC) was calculated and values greater than 0.6 were considered indicative of good reliability. Forty-eight patients (39 males, 81.2%) were enrolled. The average age at diagnosis was 12.5 years (range, 7.6 to 16.7 years). 41.7% were HLA-B27 positive and 18.8% had a positive family history for ankylosing spondylitis. 52% had involvement of the hip and 40% had radiographic evidence of sacroiliitis. Eight patients dropped out or were excluded due to protocol violation. 40 patients completed both sets of questionnaires. All subjects reported their overall health as “the same” when compared to their baseline visit. The mean BASDAI at baseline was 1.97 ± 1.90 and at 2 weeks was 1.69 ± 1.80 – the reliability was substantial; ICC = 0.74, Bland-Altman limits of agreement (LOA) = 2.4 to -2.8. The mean BASFI at baseline was 0.99 ± 1.49 and at 2 weeks was 0.75 ± 1.00 – likewise the reliability was excellent; ICC = 0.87, Bland-Altman LOA = 1.1 to -1.4. When examining individual questions from the BASDAI and BASFI, the following had the highest ICCs, respectively: “How long does your morning stiffness last from the time you wake up?” and “Doing a full day’s activities, whether it be at home or at work”, ICC = 0.88 each. Meanwhile, from the BASDAI, “How would you describe the overall level of AS neck, back or hip pain you have had?” resulted in the lowest reliability; ICC = 0.57. At this time, there is no validated disease activity score for JSpA/ERA. Both the BASDAI and BASFI showed excellent intra-rater reliability in a cohort of ERA patients. Next steps will include the measurement of the construct validity and responsiveness of these tools in JSpA/ERA in order to determine if pediatric rheumatologists can use these as validated disease activity/functional impairment measures in the clinic and in clinical research.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.019 | 0.020 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".