The Completed SDSS-IV extended Baryon Oscillation Spectroscopic Survey: exploring the halo occupation distribution model for emission line galaxies
Bibliographic record
Abstract
ABSTRACT We study the modelling of the halo occupation distribution (HOD) for the eBOSS DR16 emission line galaxies (ELGs). Motivated by previous theoretical and observational studies, we consider different physical effects that can change how ELGs populate haloes. We explore the shape of the average HOD, the fraction of satellite galaxies, their probability distribution function (PDF), and their density and velocity profiles. Our baseline HOD shape was fitted to a semi-analytical model of galaxy formation and evolution, with a decaying occupation of central ELGs at high halo masses. We consider Poisson and sub/super-Poissonian PDFs for satellite assignment. We model both Navarro–Frenk–White and particle profiles for satellite positions, also allowing for decreased concentrations. We model velocities with the virial theorem and particle velocity distributions. Additionally, we introduce a velocity bias and a net infall velocity. We study how these choices impact the clustering statistics while keeping the number density and bias fixed to that from eBOSS ELGs. The projected correlation function, wp, captures most of the effects from the PDF and satellites profile. The quadrupole, ξ2, captures most of the effects coming from the velocity profile. We find that the impact of the mean HOD shape is subdominant relative to the rest of choices. We fit the clustering of the eBOSS DR16 ELG data under different combinations of the above assumptions. The catalogues presented here have been analysed in companion papers, showing that eBOSS RSD+BAO measurements are insensitive to the details of galaxy physics considered here. These catalogues are made publicly available.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".