Validation and reliability testing of the Breast-Q latissimus dorsi questionnaire: cross-cultural adaptation and psychometric properties in a Swedish population
Bibliographic record
Abstract
BACKGROUND: The main aim of post-mastectomy breast reconstruction is to improve the patient's quality of life, which makes high-quality and validated patient-reported outcome measurements essential. None of the established instruments include evaluation of donor-site morbidity, such as impact on upper extremity and back function, when a latissimus dorsi (LD) muscle is used; and BREAST-Q LD questionnaire was therefore recently developed for this purpose. The aim of this study was to translate into Swedish and culturally adapt the BREAST-Q LD questionnaire's two subscales, appearance and function, and perform a psychometric evaluation of the subscales in a Swedish population of patients. METHODS: This was a cross-sectional study. The questionnaire was translated according to established guidelines. The questionnaires were sent to all patients operated using an LD flap between 2007 and 2017. Internal consistency was assessed using Cronbach's α. Inter-item correlations and corrected item-total correlations were calculated using the Pearson's correlation coefficient. Convergent validity was evaluated by comparing the BREAST-Q LD questionnaire to the Western Ontario Osteoarthritis of the Shoulder Index, using the Spearman correlation coefficient. Test-retest reliability was tested with intraclass correlation coefficients (ICCs), and the coefficient of variation and Bland-Altman plots were drawn. Floor and ceiling effects were calculated. Known-group validation was tested by comparing scores from the patients and from normal controls using the Mann-Whitney U-test and by calculating eta squared effect size. RESULTS: The questionnaires were sent to 176 eligible patients and 125 responded (71%). The patients had been operated a mean of 6.6 years ago, and most (92%) had previous radiation. Internal consistency was satisfactory for both subscales. The correlation coefficients between questions were r > 0.30 for all items of both scales. The corrected item-total correlation coefficient ranged from 0.62 to 0.90. As hypothesised, the function scale was correlated with the WOOS "Physical symptoms" subscale. Reliability was adequate according to the ICCs. The ceiling effect threshold for the appearance scale was reached and that for the back scale was almost reached. There were significant differences between patients and controls, in the hypothesised direction. CONCLUSIONS: The results of this study support a good internal consistency, convergent validity, test-retest reliability and known-group validation for the Swedish BREAST-Q LD questionnaire. However, it may be difficult to discriminate between patients with very mild and those with no symptoms using the appearance scale. TRIAL REGISTRATION: ClinicalTrials.Gov identifier NCT04526561.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".