135: Assessing the Accuracy of Physical Literacy Screening Tasks with the Canadian Assessment of Physical Literacy (CAPL)
Bibliographic record
Abstract
Physical literacy is a child's capacity to achieve a healthy, active lifestyle. Leaders in healthcare and allied-health (health) are important partners for identifying children with physical literacy deficits. Current physical literacy assessments are time consuming and require resources not typically available in health settings. We sought to develop and evaluate physical literacy screening tasks that would be suitable for health professionals in a variety of settings. The goal was to enable health professionals to quickly and accurately identify children in the lowest 10th percentile of physical literacy scores. Children, eight to 12 years, were recruited from recreation, education and health settings. They performed the Canadian Assessment of Physical Literacy (CAPL), a detailed assessment of a child's capacity for a physically active lifestyle. Children also answered simple questions about their own physical activity and performed eight potential physical literacy screening tasks, including one test of strength, two balance tests, one endurance test, and four motor skills assessments. Children were grouped as above/below the 10th percentile for CAPL score. Sensitivity and specificity scores for each screening task, and for combined pairs of screening tasks were calculated based on the 10th percentile groups. Study protocols were completed by 105 children (52.4% female). Sensitivity of individual test items to identify children below the 10th percentile for CAPL score ranged from 60% to 100%. Specificity for individual scores ranged from 9.6% to 87.8%. Two combination protocols were identified as having high specificity and sensitivity. Each combination contained two screening tasks: a) holding a one-leg balance test on the left leg for less than 34 seconds and a wall sit for less than 33 seconds (80.0% sensitivity, 93.3% specificity), and b) holding a one-leg balance test on the left leg for less than 40 seconds and self-reporting an activity level, as compared to their peers, below 6 out of 10 (60.0% sensitivity, 93.0% specificity). The positive predictive value of these protocols was high (98.8% and 97.9%, respectively). Two sensitive and specific physical literacy screening protocols for children eight to 12 years of age were identified for health settings. Reliability of the screening protocols and effectiveness of the screening task educational materials are on-going. Future research should evaluate the suitability of the screening tasks for children with identified disabilities/chronic illnesses, and the ability of healthcare and allied-health professionals to implement the screening tasks and utilize the results obtained for patient care.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.005 | 0.015 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.002 | 0.001 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.002 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".