Supporting data for publication: The role of the three-dimensional geometry of fault steps on event migration during fluid-induced seismic sequences.
Bibliographic record
Abstract
This repository contains the seismicity catalogues from Cahuilla, Yellowstone and West Bohemia used in the publication Roche et al., 2023 (The role of the three-dimensional geometry of fault steps on event migration during fluid-induced seismic sequences) and modified from Ross et al. (2020), Shelly et al. (2013) and Hainzl et al. (2016). The catalogues include the hypocentre location, relative time, and magnitude for non-filtered and filtered data. General information on each catalogue, filtering and modifications can be found in the associated publication. Dataset list: Cahuilla Catalogues (modified from Ross et al., 2019): Original data:File name: VR_sup_0021_Cah_All Filtered data: File name: VR_sup_0022_Cah_Filter Bohemia 2008 Catalogues (modified from Haintzl et al., 2016): Original data:File name: VR_sup_0023_Boh_08_All Filtered data: File name: VR_sup_0024_Boh_08_Filter Bohemia 2014 Catalogues (modified from Haintzl et al., 2016): Original data:File name: VR_sup_0025_Boh_14_All Filtered data: File name: VR_sup_0026_Boh_14_Filter Yellowstone Catalogs (modified from Shelly et al., 2013): Original data:File name: VR_sup_0027_Yell_14_All Filtered data: File name: VR_sup_0028_Yell_14_Filter The files are text files tab-delimited, with the following headers: Index: 1 by default Easting(m): hypocenter Easting in meters Northing(m): hypocenter Northing in meters Depth(m): hypocenter depth in meters Mw: magnitude Relative Time(s): date of the origin time in the format If you find these data useful in your research, please cite Roche et al. (2023), as well as the relevant papers Ross et al. (2020), Shelly et al. (2013) and Hainzl et al. (2016).
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.002 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.005 | 0.004 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".