Development and validation of a feline abdominal palpation model and scoring rubric
Bibliographic record
Abstract
Simulation in veterinary education enables clinical skills practice without animal use. A feline abdominal palpation model was created that allows practice in this fractious species. This study assessed the model and rubric using a validation framework of content evidence, internal structure and relationship with level of training. Content Evidence: Veterinarians accepted this model as a helpful training tool for students (median=4 on five-point Likert scale). Internal Structure Evidence: G-coefficients were low for first- and second-year students (0.28 and 0.23), but were acceptable for veterinarians (0.61). Internal consistency values (0.24, 0.42 and 0.67) followed a similar pattern. Thus, scores were more reliable for veterinarians than for the students. Evidence of Relationship with Level of Training: Although level of training impacted reliability, its effect on performance scores was inconsistent. Analysis of variance (ANOVA) identified no differences among the groups of students and veterinarians. However, effect size between first- and third-year students was medium to large (0.62). Effect sizes between the veterinarians and student groups were small. Although the model and rubric appeared valid for experts, modifications would be necessary to generate reliable scores for students. These results allow greater understanding of the needs of students utilising a low-fidelity model.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".