Using Crash Data to Develop Simulator Scenarios for Assessing Novice Driver Performance
Bibliographic record
Abstract
Teenage drivers are at their highest crash risk in their first 6 months or first 1,000 mi of driving. Driver training, adult-supervised practice driving, and other interventions are aimed at improving driving performance in novice drivers. Previous driver training programs have enumerated thousands of scenarios, with each scenario requiring one or more skills. Although there is general agreement about the broad set of skills needed to become a competent driver, there is no consensus set of scenarios and skills to assess whether novice drivers are likely to crash or to assess the effects of novice driver training programs on the likelihood of a crash. The authors propose that a much narrower, common set of scenarios can be used to focus on the high-risk crashes of young drivers. Until recently, it was not possible to identify the detailed set of scenarios that were specific to high-risk crashes. However, an integration of police crash reports from previous research, a number of critical simulator studies, and a nationally representative database of serious teen crashes (the National Motor Vehicle Crash Causation Survey) now make identification of these scenarios possible. In this paper, the authors propose this novel approach and discuss how to create a common set of simulated scenarios and skills to assess novice driver performance and the effects of training and interventions as they relate to high-risk crashes.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.005 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.002 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".