Determining the fraction of reddened quasars in COSMOS with multiple selection techniques from X-ray to radio wavelengths
Bibliographic record
Abstract
The sub-population of quasars reddened by intrinsic or intervening clouds of dust are known to be underrepresented in optical quasar surveys. By defining a complete parent sample of the brightest and spatially unresolved quasars in the COSMOS field, we quantify to which extent this sub-population is fundamental to our understanding of the true population of quasars. By using the available multiwavelength data of various surveys in the COSMOS field, we built a parent sample of 33 quasars brighter than J = 20 mag, identified by reliable X-ray to radio wavelength selection techniques. Spectroscopic follow-up with the NOT/ALFOSC was carried out for four candidate quasars that had not been targeted previously to obtain a 100% redshift completeness of the sample. The population of high AV quasars (HAQs), a specific sub-population of quasars selected from optical/near-infrared photometry, some of which were shown to be missed in large optical surveys such as SDSS, is found to contribute 21%+9-5 of the parent sample. The full population of bright spatially unresolved quasars represented by our parent sample consists of 39%+9-8 reddened quasars defined by having AV > 0.1, and 21%+9-5 of the sample having E(B−V) > 0.1 assuming the extinction curve of the Small Magellanic Cloud. We show that the HAQ selection works well for selecting reddened quasars, but some are missed because their optical spectra are too blue to pass the g−r color cut in the HAQ selection. This is either due to a low degree of dust reddening or anomalous spectra. We find that the fraction of quasars with contributing light from the host galaxy, causing observed extended spatial morphology, is most dominant at z ≲ 1. At higher redshifts the population of spatially unresolved quasars selected by our parent sample is found to be representative of the full population of bright active galactic nuclei at J< 20 mag. This work quantifies the bias against reddened quasars in studies that are based solely on optical surveys.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".