The creation and validation of a short form of the Neurogenic Bladder Symptom Score
Bibliographic record
Abstract
AIM: To develop a short form (SF) of the 24-item Neurogenic Bladder Symptom Score (NBSS). METHODS: We used three previously published datasets. First, we selected the most responsive questions within each of the domains. Internal validity of the NBSS-SF was assessed using Cronbach's α. External validity was assessed by evaluating hypothesized relationships with other questionnaires and testing correlations with the full NBSS domains. Test-retest reliability of the NBSS-SF domains was determined using an intraclass coefficient (ICC). RESULTS: Using data from a prior responsiveness study, we selected questions for the NBSS-SF from the incontinence domain (three), storage/voiding domain (three), consequences domain (two); these would make up the NBSS-SF. We used the original NBSS validation cohort of 230 patients with multiple sclerosis (MS), spinal cord injury (SCI), or spina bifida, and found the Cronbach's α was .76 for the NBSS-SF; the external validity was high, with correlations between specific NBSS-SF domains/total scores and the Qualiveen-SF, ICIQ, and AUASS generally similar to those seen with the NBSS. Correlations between the NBSS-SF domains and the full NBSS domains were high. The NBSS-SF ICC in a subset of 120 patients was 0.84. The NBSS-SF performed similarly in two additional independent datasets. CONCLUSIONS: The total score of the NBSS-SF has appropriate validity, reliability, and could be used instead of the full NBSS to minimize the assessment burden. The full NBSS may be better suited if the primary focus of the study is on neurogenic bladder symptoms, or if individual NBSS domains are of interest.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".