MétaCan
Menu
Back to cohort
Record W3048979592 · doi:10.1080/01441647.2020.1806943

Crowdsourced data for bicycling research and practice

2020· article· en· W3048979592 on OpenAlexafffund
Trisalyn Nelson, Colin Ferster, Karen Laberee, Daniel Fuller, Meghan Winters

Bibliographic record

VenueTransport Reviews · 2020
Typearticle
Languageen
FieldSocial Sciences
TopicUrban Transport and Accessibility
Canadian institutionsMemorial University of NewfoundlandSimon Fraser UniversityUniversity of Victoria
FundersMichael Smith Health Research BCPublic Health Agency of CanadaCanada Research ChairsArizona State University
KeywordsCrowdsourcingTransport engineeringGlobal Positioning SystemTRIPS architectureCitizen scienceTraffic congestionSocial mediaOpen dataPoison controlComputer scienceData scienceEngineeringWorld Wide Web

Abstract

fetched live from OpenAlex

Cities are promoting bicycling for transportation as an antidote to increased traffic congestion, obesity and related health issues, and air pollution. However, both research and practice have been stalled by lack of data on bicycling volumes, safety, infrastructure, and public attitudes. New technologies such as GPS-enabled smartphones, crowdsourcing tools, and social media are changing the potential sources for bicycling data. However, many of the developments are coming from data science and it can be difficult evaluate the strengths and limitations of crowdsourced data. In this narrative review we provide an overview and critique of crowdsourced data that are being used to fill gaps and advance bicycling behaviour and safety knowledge. We assess crowdsourced data used to map ridership (fitness, bike share, and GPS/accelerometer data), assess safety (web-map tools), map infrastructure (OpenStreetMap), and track attitudes (social media). For each category of data, we discuss the challenges and opportunities they offer for researchers and practitioners. Fitness app data can be used to model spatial variation in bicycling ridership volumes, and GPS/accelerometer data offer new potential to characterise route choice and origin-destination of bicycling trips; however, working with these data requires a high level of training in data science. New sources of safety and near miss data can be used to address underreporting and increase predictive capacity but require grassroots promotion and are often best used when combined with official reports. Crowdsourced bicycling infrastructure data can be timely and facilitate comparisons across multiple cities; however, such data must be assessed for consistency in route type labels. Using social media, it is possible to track reactions to bicycle policy and infrastructure changes, yet linking attitudes expressed on social media platforms with broader populations is a challenge. New data present opportunities for improving our understanding of bicycling and supporting decision making towards transportation options that are healthy and safe for all. However, there are challenges, such as who has data access and how data crowdsourced tools are funded, protection of individual privacy, representativeness of data and impact of biased data on equity in decision making, and stakeholder capacity to use data given the requirement for advanced data science skills. If cities are to benefit from these new data, methodological developments and tools and training for end-users will need to track with the momentum of crowdsourced data.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.126
metaresearch head score (Gemma)0.344
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: none
GenreCandidate signal: Review · Consensus signal: none
Teacher disagreement score0.126
Threshold uncertainty score0.667

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.1260.344
Meta-epidemiology (narrow)0.0020.001
Meta-epidemiology (broad)0.0030.003
Bibliometrics0.0130.015
Science and technology studies0.0050.009
Scholarly communication0.0140.015
Open science0.0070.021
Research integrity0.0060.005
Insufficient payload (model declined to judge)0.0200.008

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.597
GPT teacher head0.538
Teacher spread0.059 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designNot applicable
Domainnot available
GenreReview

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations82
Published2020
Admission routes2
Has abstractyes

Explore more

Same venueTransport ReviewsSame topicUrban Transport and AccessibilityFrench-language works237,207