Using Host Galaxy Photometric Redshifts to Improve Cosmological Constraints with Type Ia Supernovae in the LSST Era
Bibliographic record
Abstract
Abstract We perform a rigorous cosmology analysis on simulated Type Ia supernovae (SNe Ia) and evaluate the improvement from including photometric host galaxy redshifts compared to using only the “z spec” subset with spectroscopic redshifts from the host or SN. We use the Deep Drilling Fields (∼50 deg2) from the Photometric LSST Astronomical Time-Series Classification Challenge (PLAsTiCC) in combination with a low-z sample based on Data Challenge2. The analysis includes light-curve fitting to standardize the SN brightness, a high-statistics simulation to obtain a bias-corrected Hubble diagram, a statistical+systematics covariance matrix including calibration and photo-z uncertainties, and cosmology fitting with a prior from the cosmic microwave background. Compared to using the z spec subset, including events with SN+host photo-z results in (i) more precise distances for z > 0.5, (ii) a Hubble diagram that extends 0.3 further in redshift, and (iii) a 50% increase in the Dark Energy Task Force figure of merit (FoM) based on the w 0 w a CDM model. Analyzing 25 simulated data samples, the average bias on w 0 and w a is consistent with zero. The host photo-z systematic of 0.01 reduces FoM by only 2% because (i) most z < 0.5 events are in the z spec subset, (ii) the combined SN+host photo-z has ×2 smaller bias, and (iii) the anticorrelation between fitted redshift and color self-corrects distance errors. To prepare for analyzing real data, the next SN Ia cosmology analysis with photo-zs should include non–SN Ia contamination and host galaxy misassociations.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.005 | 0.013 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".