Oxford Shoulder Instability Score: cross-cultural adaptation into Spanish and analysis of its methodological quality
Bibliographic record
Abstract
BACKGROUND: Oxford Shoulder Instability Score (OSIS) is a patient reported outcome measure designed specifically to assess functional difficulties resulting from shoulder instability. The main aim of this study was to cross-culturally adapt the OSIS to Spanish. Secondary, it aimed to analyse its methodological quality. METHODS: A cross-cultural adaptation to Spanish has been carried out following the recommendations of COnsensus-based Standards for the selection of health Measurement Instruments (COSMIN) with a sample of 167 cases of shoulder instability. Inclusion criteria were: subjects with instability symptoms in at least one shoulder with or without clinical diagnosis; aged between 18 and 60. The following psychometric properties were evaluated: validity (construct, internal and external), reliability (internal consistency, test-retest and measurement error), discriminant ability and feasibility. The methodological quality was addressed with Quality Assessment of Diagnostic Accuracy Studies-2 (QUADAS-2) and COSMIN Risk of Bias checklist (COSMIN RoB). RESULTS: The construct validity obtained an OSIS correlation with the Simple Shoulder Test of r = 0.636 and with the Western Ontario Shoulder Instability of r = 0.80. The unidimensionality of the OSIS was confirmed through second-order factor analysis. Cronbach's alpha and intraclass correlation coefficient were both 0.93 (IC95%: 0.91-0.94). Standard error of measurement was 0.70, and the percentage of error and smallest detectable change were 1.46% and 1.94, respectively. No floor or ceiling effects were found. Assessing feasibility of OSIS, the participants answered all questions, had no questions and completion time was: mean 2 min 30 s; SD ± 1 min. Regarding methodological quality, the study showed low risk of bias in the areas of patient selection, index test, reference standard and flow and timing, as well as low concern regarding applicability in the domains of patient selection, index test and reference standard according to QUADAS-2. In relation to COSMIN RoB, all psychometric properties, except for content validity, obtained very good results. CONCLUSIONS: The Spanish version of OSIS offers valid, reliable and feasible functional outcome measures for Spanish-speaking subjects with shoulder instability, based on its psychometric properties and its methodological quality.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.075 | 0.101 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.003 |
| Bibliometrics | 0.008 | 0.007 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.003 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".