MétaCan
Menu
Back to cohort
Record W2973246940 · doi:10.11575/prism/37058

Cluster Analysis of Long-Distance Person Travel in Alberta

2019· dissertation· en· W2973246940 on OpenAlexaboutno aff
Mina Hassanvand

Bibliographic record

VenueOpen MIND · 2019
Typedissertation
Languageen
FieldSocial Sciences
TopicHuman Mobility and Location-Based Analysis
Canadian institutionsnot available
Fundersnot available
KeywordsCluster (spacecraft)GeographyTransport engineeringComputer scienceEngineering

Abstract

fetched live from OpenAlex

Transportation is the movement of goods and people throughout a network of links and nodes and consists of short-distance SD (intra-city) and long-distance LD (inter-city) trips. A component of the latter which is responsible for a relatively large portion of total kilometres travelled – that is long-distance person trips LDPT not including LD trucking – has yet to be profoundly investigated, measured, and modelled to become manageable and benefit relevant policies and funding allocations. In LDPT modelling, the traditional methods of LD trips representation are those used for SD trips modelling in which some explanatory factors such trip purpose are selected for LDPT segmentation as a means to distinguish among sectors of such trips. This in turn has resulted in the development of LDPT models that fall short of correctly representing such trips. In this regard, the work described herein hypothesizes and demonstrates how LD trips are not merely longer versions of SD trips, but their relevant datasets possess a natural structure and inherent groupings or clusters of trips that needs to be deeply investigated using appropriate methods (cluster analysis used in computer science studies for network related data) in order to contribute to much needed refinements of LDPT modelling exercises. The method along with several novel avenues of input selection, validity, and significance measurements, revealed ten heterogeneous clusters of LDPT to exist in the 2010 Alberta to Alberta trips in the 2010 Travel Survey of Residents of Canada TSRC dataset. The clusters exhibiting a mix of trip purpose, person, and activity features, with proportions of total person trips shown in brackets are: short economical getaways (31%), same-day shopping (22%), personal business (12%), visiting friends and relatives (10%), business/casino trips (9%), young adult team sport players (4%), same-day trips of snow/festival loving young families with kids (4%), costly cottage trips (3%), high educated multiple city visitors (3%), and seniors with medical appointments (2%). The existence of clusters which were detected through the approach proposed herein carries a fundamental meaning showing there exist different segments of LD travel behaviour in reality which need to be studied if one is to accurately model them.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.001
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesInsufficient payload (model declined to judge)
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Qualitative · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.763
Threshold uncertainty score0.989

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0010.000
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0010.000
Bibliometrics0.0000.002
Science and technology studies0.0000.000
Scholarly communication0.0000.000
Open science0.0010.000
Research integrity0.0000.000
Insufficient payload (model declined to judge)0.0120.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.033
GPT teacher head0.352
Teacher spread0.319 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

Study designQualitative
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2019
Admission routes1
Has abstractyes

Explore more

Same venueOpen MINDSame topicHuman Mobility and Location-Based AnalysisFrench-language works237,207