Hydroxyl Lines and Moonlight: A High Spectral Resolution Investigation of Near-infrared Skylines from Maunakea to Guide Near-infrared Spectroscopic Surveys
Bibliographic record
Abstract
Abstract Subtracting the changing sky contribution from the near-infrared (NIR) spectra of faint astronomical objects is challenging and crucial to a wide range of science cases such as estimating the velocity dispersions of dwarf galaxies, studying the gas dynamics in faint galaxies, measuring accurate redshifts, and any spectroscopic study of faint targets. Since the sky background varies with time and location, NIR spectral observations, especially those employing fiber spectrometers and targeting extended sources, require frequent sky-only observations for calibration. However, sky subtraction can be optimized with sufficient a priori knowledge of the sky's variability. In this work, we explore how to optimize sky subtraction by analyzing 1075 high-resolution NIR spectra from the Canada–France–Hawaii Telescope's SPIRou on Maunakea, and we estimate the variability of 481 hydroxyl (OH) lines. These spectra were collected during two sets of three nights dedicated to obtaining sky observations every 5.5 minutes. During the first set, we observed how the Moon affects the NIR, which has not been accurately measured at these wavelengths. We suggest accounting for the Moon contribution at separation distances less than 10° when (1) reconstructing the sky using principal component analysis, (2) observing targets at YJHK magnitudes fainter than ∼15, and (3) attempting a sky subtraction better than 1%. We also identified 126 spectral doublets, or OH lines that split into at least two components, at SPIRou's resolution. In addition, we used Lomb–Scargle periodograms and Gaussian process regression to estimate that most OH lines vary on similar timescales, which provides a valuable input for IR spectroscopic survey strategies. The data ( https://zenodo.org/records/13363061 ) and code ( https://github.com/FDauphin/spirou-sky-subtraction ) developed for this study are publicly available.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".