<i>Euclid</i> preparation
Bibliographic record
Abstract
We present two extensive sets of 3500+1000 simulations of dark matter haloes on the past light cone and two corresponding sets of simulated (mock) galaxy catalogues that represent the spectroscopic sample of Euclid . The simulations were produced with the latest version of the code Pinocchio and provide the largest public set of simulated skies. The mock galaxy catalogues were obtained by populating haloes with galaxies using an halo occupation distribution (HOD) model extracted from the Flagship galaxy catalogue provided by Euclid Collaboration. The Geppetto set of 3500 simulated skies was obtained by tiling a 1.2 h −1 Gpc box to cover a light cone whose sky footprint is a circle with a radius of 30° for an area of 2763 deg 2 and a minimum halo mass of 1.5 × 10 11 h −1 M ⊙ . The relatively small size of the box means that this set is unsuitable for measuring very large scales. The EuclidLargeBox set consists of 1000 simulations of 3.38 h −1 Gpc and has the same mass resolution and a footprint that covers half of the sky. It excludes the Milky Way zone of avoidance. From this, we produced a set of 1000 EuclidLargeMocks on the 30° radius footprint, whose comoving volume is fully contained in the simulation box. We validated the two sets of catalogues by analysing number densities, power spectra, and two-point correlation functions to show that the Flagship spectroscopic catalogue is consistent with being one of the realisations of the simulated sets. We noted small deviations, however, that are limited to the quadrupole at k > 0.2 h Mpc −1 . We infer the cosmological parameters from these catalogues and demonstrate that using one realisation of EuclidLargeMocks in place of the Flagship mock produces the same posteriors to within the expected shift given by the sample variance. These simulated skies will be used for the galaxy clustering analysis of the Euclid Data Release 1 (DR1), and an even larger set of simulations is planned for the next releases.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.001 | 0.003 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".