Exercise-Induced Tendon and Bone Injury in Recreational Runners: A Test-Retest Reliability Study
Bibliographic record
Abstract
BACKGROUND: Long-distance runners are prone to injuries including Achilles tendinopathy and medial tibial stress syndrome. We have developed an Internet comprehensive self-report questionnaire examining the medical history, injury history, and running habits of adult recreational runners. OBJECTIVE: The objective of the study was to evaluate two alternative forms of test-retest reliability of a comprehensive self-report Internet questionnaire retrospectively examining the medical history, injury history, and running habits among a sample of adult recreational runners. This will contribute to the broad aims of a wider study investigating genetics and running injury. METHODS: Invitations to complete an Internet questionnaire were sent by email to a convenience pilot population (test group 1). Inclusion criteria required participants to be a recreational runner age 18 or over, who ran over 15 km per week on a consistent basis. The survey questions addressed regular running habits and any injuries (including signs, symptoms, and diagnosis) of the lower limbs that resulted in discontinuation of running for a period of 2 consecutive weeks or more, within the last 2 years. Questions also addressed general health, age, sex, height, weight, and ethnic background. Participants were then asked to repeat the survey using the Internet platform again after 10-14 days. Following analysis of test group 1, we soft-launched the survey to a larger population (test group 2), through a local running club of 900 members via email platform. The same inclusion criteria applied, however, participants were asked to complete a repeat of the survey by telephone interview after 7-10 days. Selected key questions, important to clarify inclusion or exclusion from the wider genetics study, were selected to evaluate test-retest reliability. Reliability was quantified using the kappa coefficient for categorical data. RESULTS: In response to the invitation, 28 participants accessed the survey from test group 1, 23 completed the Internet survey on the first occasion, and 20 completed the Internet retest within 10-21 days. Test-retest reliability scored moderate to almost perfect (kappa=.41 to .99) for 19/19 of the key questions analyzed. Following the invitation, 122 participants accessed the survey from test group 2, 101 completed the Internet survey on the first occasion, and 50 were randomly selected and contacted by email inviting them to repeat the survey by telephone interview. There were 33 participants that consented to the telephone interview and 30 completed the questionnaire within 7-10 days. Test-retest reliability scored moderate to almost perfect for 18/19 (kappa=.41 to .99) and slight for 1/19 of the key questions analyzed.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".