Joint Confidence Region Approach to Ranking Hotspot Locations Considering Uncertainty in Expected Risk Estimates
Notice bibliographique
Résumé
Network screening or crash hotspot identification is an essential task of all road safety improvement programs. The most common approach to network screening is to use statistical models to predict the expected risk at the locations of interest and then rank them accordingly. The predicted risk used for ranking is mostly in the form of point estimates, without any consideration of the inherent uncertainty with the estimates, which could lead to identifying a wrong list of crash hotspots. This study aims to fill this research gap by employing a frequentist approach to finding a joint confidence region of risk for ranking locations and identification of hotspots. A case study on three-legged minor approach stop-controlled intersections in Kitchener, Ontario, is conducted to illustrate the proposed approach. Crash risk is modeled using a combination of a hierarchical full Bayesian negative binomial model and a multinomial logit model, which are then used to estimate the 95% confidence interval of the expected risk. For each location, the confidence region of rankings is obtained on the basis of the expected risk estimates. The results show that considering uncertainty in the crash hotspot identification process can lead to varied ranking positions for each location. In fact, considering uncertainty, the true value of the estimated crash risk is unknown. By quantifying uncertainty, it can be concluded that the true value of the estimated risk follows a distribution with different probabilities. As a result, consideration of uncertainty in the road safety analysis may help to identify hotspots more accurately.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,019 | 0,075 |
| Méta-épidémiologie (sens strict) | 0,002 | 0,002 |
| Méta-épidémiologie (sens large) | 0,003 | 0,003 |
| Bibliométrie | 0,006 | 0,004 |
| Études des sciences et des technologies | 0,001 | 0,003 |
| Communication savante | 0,004 | 0,004 |
| Science ouverte | 0,005 | 0,003 |
| Intégrité de la recherche | 0,003 | 0,003 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,004 | 0,001 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».