Deep Near-Infrared Survey Towards the W40 and Serpens South Region in the Aquila Rift: A Comprehensive Catalogue of Young Stellar Objects
Bibliographic record
Abstract
Active star-forming regions are excellent laboratories for studying the origins and evolution of young stellar object (YSO) clustering. The W40–Serpens South region is such a region, and we compile a large near- and mid-infrared catalogue of point sources in it, based on deep near-infrared observations of Canada-France-Hawaii Telescope (CFHT) in combination with Two Micron All Sky Survey (2MASS), UKIRT Infrared Deep Sky Survey (UKIDSS), and Spitzer catalogues. From this catalogue, we identify 832 YSOs, and classify 15, 135, 647, and 35 of them to be deeply embedded sources, Class I YSOs, Class II YSOs, and transition disc sources, respectively. In general, these YSOs are well correlated with the filamentary structures of molecular clouds, especially the deeply embedded sources and the Class I YSOs. The W40 central region is dominated by Class II YSOs, but in the Serpens South region, half of the YSOs are Class I. We further generate a minimum spanning tree (MST) for all the YSOs. Around the W40 cluster, there are eight prominent MST branches that may trace the vestigial molecular gas filaments that once fed gas to the central natal gas clump. Of the eight, only two now include detectable filamentary gas in Herschel data and corresponding Class I YSOs, while the other six are populated exclusively with Class II YSOs. Four MST branches overlap with the Serpens South main filament, and where they intersect, molecular gas ‘hubs’ and more Class I YSOs are found. Our results imply a mixture of YSO distributions composed of both primordial and somewhat evolved YSOs in this star-forming region.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.002 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".