108 An Introductory Systematic Review of the Vancouver Scar Scale: Versions, Validations and Utilizations
Bibliographic record
Abstract
Abstract Introduction Since its development in 1990, the Vancouver Scar Scale (VSS) has been validated and modified in order to objectively rate the appearance of a burn scar. The VSS subjectively assesses the pigmentation, vascularity, pliability, and height of the burn scar and was the first scale that attempted to capture these components of maturing burn scars. In 2024, this tool is still considered one of the most common evaluation methods for scars despite low inter-rater reliability, inconsistent validity, and multiple modifications. A systematic review was completed to assess VSS and its modifications, as well as current trends in its utilization. Methods A PRISMA-compliant qualitative systematic review was conducted on studies published since 1990 making reference to the VSS. One thousand twenty seven studies were provided as a result of the search. Due to the volume of abstracts found, this introductory review was further limited to then include only the 107 studies published between January 1, 2024 and September 30, 2024. Results Of the 107 studies reviewed, only 34 were excluded due to lack of full text availability, two of which were not available in English. Majority of studies were assessing non-burn scars, did not specify the parameters of the VSS, nor provide any explanation of the assessment. Only 5 of the included 73 studies comprehensively interpreted the VSS outcomes and only 3 acknowledged the weakness of the VSS as a comprehensive scar assessment. Several articles added assessment components such as pain and pruritus while others adjusted the scoring system for the tool. Nearly half of the publications reviewed did not provide a reference to indicate which version of the VSS was being utilized. Majority of the included studies were published in peer-reviewed journals which fall in the domains of Plastic Surgery, Dermatology, and Orthopedics. Conclusions The VSS has been validated for use with burn scars, yet its current utilization in published literature falls outside of burn-specific populations, while its limitations and poor inter-rater reliability are not fully acknowledged in these publications. Furthermore, the insufficient documentation referencing the specific version of the VSS utilized brings forth concerns for validity and credibility of some published studies. Further review is required to fully assess the progression from the original validation of the VSS to its present day general utilization. Applicability of Research to Practice To the extent that the VSS is being utilized in practice, research, and publications, this study reinforces the need for further research to validate the credibility of the VSS and strengthen the impact of associated publications Funding for the Study No external funding was provided
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.032 | 0.145 |
| Meta-epidemiology (narrow) | 0.002 | 0.002 |
| Meta-epidemiology (broad) | 0.007 | 0.008 |
| Bibliometrics | 0.019 | 0.018 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.004 | 0.004 |
| Open science | 0.002 | 0.003 |
| Research integrity | 0.003 | 0.002 |
| Insufficient payload (model declined to judge) | 0.009 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".