MétaCan
Menu
Back to cohort
Record W4403750696 · doi:10.1126/sciadv.adn3268

Architectural styles of curiosity in global Wikipedia mobile app readership

2024· article· en· W4403750696 on OpenAlexaff
Dale Zhou, Shubhankar P. Patankar, David M. Lydon‐Staley, Perry Zurn, Martin Gerlach, Danielle S. Bassett

Bibliographic record

VenueScience Advances · 2024
Typearticle
Languageen
FieldPsychology
TopicPsychological and Educational Research Studies
Canadian institutionsMcGill UniversityMontreal Neurological Institute and Hospital
FundersNational Institute on Drug AbuseFondation pour la Recherche MédicaleGeorge E. Hewitt Foundation for Medical Research
KeywordsCuriosityAudience measurementMobile appsComputer scienceWorld Wide WebMultimediaData sciencePsychologyAdvertisingNeuroscienceBusiness

Abstract

fetched live from OpenAlex

Intrinsically motivated information seeking is an expression of curiosity believed to be central to human nature. However, most curiosity research relies on small, Western convenience samples. Here, we analyze a naturalistic population of 482,760 readers using Wikipedia's mobile app in 14 languages from 50 countries or territories. By measuring the structure of knowledge networks constructed by readers weaving a thread through articles in Wikipedia, we replicate two styles of curiosity previously identified in laboratory studies: the nomadic "busybody" and the targeted "hunter." Further, we find evidence for another style-the "dancer"-which was previously predicted by a historico-philosophical examination of texts over two millennia and is characterized by creative modes of knowledge production. We identify associations, globally, between the structure of knowledge networks and population-level indicators of spatial navigation, education, mood, well-being, and inequality. These results advance our understanding of Wikipedia's global readership and demonstrate how cultural and geographical properties of the digital environment relate to different styles of curiosity.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.001
metaresearch head score (Gemma)0.006
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Observational · Consensus signal: Observational
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.003
Threshold uncertainty score0.007

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0010.006
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0020.001
Science and technology studies0.0010.001
Scholarly communication0.0020.002
Open science0.0000.001
Research integrity0.0000.000
Insufficient payload (model declined to judge)0.0020.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.058
GPT teacher head0.459
Teacher spread0.401 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designObservational
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations11
Published2024
Admission routes1
Has abstractyes

Explore more

Same venueScience AdvancesSame topicPsychological and Educational Research StudiesFrench-language works237,207