THE LAST OF<i>FIRST</i>: THE FINAL CATALOG AND SOURCE IDENTIFICATIONS
Bibliographic record
Abstract
The FIRST survey, begun over 20 years ago, provides the definitive high-resolution map of the radio sky. This Very Large Telescope (VLA) survey reaches a detection sensitivity of 1 mJy at 20 cm over a final footprint of 10,575 deg 2 that is largely coincident with the Sloan Digital Sky Survey (SDSS) area. Both the images and a catalog containing 946,432 sources are available through the FIRST Web site ( http://sundog.stsci.edu ). We record here the authoritative survey history, including hardware and software changes that affect the catalog's reliability and completeness. In particular, we use recent observations taken with the JVLA to test various aspects of the survey data (astrometry, CLEAN bias, and the flux density scale). We describe a new, sophisticated algorithm for flagging potential sidelobes in this snapshot survey, and show that fewer than 10% of the cataloged objects are likely sidelobes, and that these are heavily concentrated at low flux densities and in the vicinity of bright sources, as expected. We also report a comparison of the survey with the NRAO VLA Sky Survey (NVSS), as well as a match of the FIRST catalog to the SDSS and Two Micron Sky Survey (2MASS) sky surveys. The NVSS match shows very good consistency in flux density scale and astrometry between the two surveys. The matches with 2MASS and SDSS indicate a systematic ∼10–20 mas astrometric error with respect to the optical reference frame in all VLA data that has disappeared with the advent of the JVLA. We demonstrate strikingly different behavior between the radio matches to stellar objects and to galaxies in the optical and IR surveys reflecting the different radio populations present over the flux density range 1–1000 mJy. As the radio flux density declines, stellar counterparts (quasars) get redder and fainter, while galaxies get brighter and have colors that initially redden but then turn bluer near the FIRST detection limit. Implications for future radio sky surveys are also briefly discussed. In particular, we show that for radio source identification at faint optical magnitudes, high angular resolution observations are essential, and cannot be sacrificed in exchange for high signal-to-noise data. The value of a JVLA survey as a complement to Square Kilometer Array precursor surveys is briefly discussed.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.006 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.007 | 0.006 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.005 | 0.003 |
| Open science | 0.001 | 0.002 |
| Research integrity | 0.001 | 0.002 |
| Insufficient payload (model declined to judge) | 0.080 | 0.094 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".