MétaCan
Menu
Back to cohort
Record W7117887987 · doi:10.17605/osf.io/e8wj9

Intuitive and Reflective Foundations of Free Will and Scientific Determinism Exp. 2b- Canada

2025· other· W7117887987 on OpenAlexaboutno aff
Berke Aydaş

Bibliographic record

VenueOpen Science Framework · 2025
Typeother
Language
Field
Topic
Canadian institutionsnot available
Fundersnot available
KeywordsDeterminismExperimental scienceReflection (computer programming)Free willData collectionStatistical hypothesis testingRobustness (evolution)

Abstract

fetched live from OpenAlex

In this project, we attempt to replicate our previous study (see OSF link: https://osf.io/p7gnz/?view_only=4f5d2f2d795642f0957e1685e4a5342e). The experiment in the previous study yielded mixed evidence regarding the relationship between reflective cognitive style and scientific determinism. Contrary to our initial hypothesis that reflection would increase the endorsement of scientific determinism while decreasing the endorsement of free will, the research showed that reflection decreased support for both free will and scientific determinism. This unexpected finding allowed us to propose a novel hypothesis called the reflective doubt hypothesis based on Yılmaz and Isler’s previous finding (2019), showing that reflection increased belief in God for non-believers, but it tended to decrease it among believers. Yılmaz and Isler (2019) argue that reflection leads to an increase in doubt in one’s initial or intuitive judgments. Hence, we assume that the decrease in scientific determinism when using reflection could be a result of doubt. Hence, we aim to confirm this unexpected finding with improved methodology. This is Phase Two of our study, where we are collecting additional data from Canada. In the initial phase, data collected from Turkey did not reach the expected statistical power, resulting in an underpowered study. To address this issue, we are expanding our data collection to Canada, a WEIRD (Western, Educated, Industrialized, Rich, Democratic) country compared to Turkey. This phase aims to strengthen the study’s robustness by increasing sample size and cultural diversity. Importantly, we are not introducing any new confirmatory hypotheses; our research objectives and hypotheses remain consistent with the original phase.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.005
metaresearch head score (Gemma)0.014
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch, Meta-epidemiology (narrow), Science and technology studies, Scholarly communication, Open science, Insufficient payload (model declined to judge)
Consensus categoriesScience and technology studies, Open science
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Theoretical or conceptual · Consensus signal: none
GenreCandidate signal: Other · Consensus signal: none
Teacher disagreement score0.678
Threshold uncertainty score0.999

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0050.014
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0020.000
Bibliometrics0.0020.011
Science and technology studies0.0040.026
Scholarly communication0.0060.003
Open science0.0090.012
Research integrity0.0010.001
Insufficient payload (model declined to judge)0.0030.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.023
GPT teacher head0.348
Teacher spread0.325 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; both teacher heads agree on what is shown here.

Study designTheoretical or conceptual
Domainnot available
GenreOther

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2025
Admission routes1
Has abstractyes

Explore more

Same venueOpen Science FrameworkFrench-language works237,207