Intra-rater reliability of the bath ankylosing spondylitis disease activity index (BASDAI) and the bath ankylosing spondylitis functional index (BASFI) in children with spondyloarthritis
Bibliographic record
Abstract
Patients diagnosed with ERA (ILAR criteria) and followed at The Hospital for Sick Children Spondyloarthritis Clinic were included in this study. Patients were excluded if they were unable to understand, speak or write English, if they were less than 6 years of age or greater than 18 years of age. Prospective subjects were consecutively enrolled ( June 2009 to June 2010) and, the patient and/or one parent completed the BASDAI and BASFI at baseline and 2 weeks later (a period during which little change is expected). Intra-class correlation coefficient (ICC) was calculated and values greater than 0.6 were considered indicative of good reliability. Forty-eight patients (39 males, 81.2%) were enrolled. The average age at diagnosis was 12.5 years (range, 7.6 to 16.7 years). 41.7% were HLA-B27 positive and 18.8% had a positive family history for ankylosing spondylitis. 52% had involvement of the hip and 40% had radiographic evidence of sacroiliitis. Eight patients dropped out or were excluded due to protocol violation. 40 patients completed both sets of questionnaires. All subjects reported their overall health as “the same” when compared to their baseline visit. The mean BASDAI at baseline was 1.97 ± 1.90 and at 2 weeks was 1.69 ± 1.80 – the reliability was substantial; ICC = 0.74, Bland-Altman limits of agreement (LOA) = 2.4 to -2.8. The mean BASFI at baseline was 0.99 ± 1.49 and at 2 weeks was 0.75 ± 1.00 – likewise the reliability was excellent; ICC = 0.87, Bland-Altman LOA = 1.1 to -1.4. When examining individual questions from the BASDAI and BASFI, the following had the highest ICCs, respectively: “How long does your morning stiffness last from the time you wake up?” and “Doing a full day’s activities, whether it be at home or at work”, ICC = 0.88 each. Meanwhile, from the BASDAI, “How would you describe the overall level of AS neck, back or hip pain you have had?” resulted in the lowest reliability; ICC = 0.57. At this time, there is no validated disease activity score for JSpA/ERA. Both the BASDAI and BASFI showed excellent intra-rater reliability in a cohort of ERA patients. Next steps will include the measurement of the construct validity and responsiveness of these tools in JSpA/ERA in order to determine if pediatric rheumatologists can use these as validated disease activity/functional impairment measures in the clinic and in clinical research.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".