The Pristine survey – I. Mining the Galaxy for the most metal-poor stars
Notice bibliographique
Résumé
We present the Pristine survey, a new narrow-band photometric survey focused on the metallicity-sensitive Ca H&K lines and conducted in the Northern hemisphere with the wide-field imager MegaCam on the Canada–France–Hawaii Telescope. This paper reviews our overall survey strategy and discusses the data processing and metallicity calibration. Additionally we review the application of these data to the main aims of the survey, which are to gather a large sample of the most metal-poor stars in the Galaxy, to further characterize the faintest Milky Way satellites, and to map the (metal-poor) substructure in the Galactic halo. The current Pristine footprint comprises over 1000 deg2 in the Galactic halo ranging from b ∼ 30° to ∼78° and covers many known stellar substructures. We demonstrate that, for Sloan Digital Sky Survey (SDSS) stellar objects, we can calibrate the photometry at the 0.02-mag level. The comparison with existing spectroscopic metallicities from SDSS/Sloan Extension for Galactic Understanding and Exploration (SEGUE) and Large Sky Area Multi-Object Fiber Spectroscopic Telescope shows that, when combined with SDSS broad-band g and i photometry, we can use the CaHK photometry to infer photometric metallicities with an accuracy of ∼0.2 dex from [Fe/H] = −0.5 down to the extremely metal-poor regime ([Fe/H] < −3.0). After the removal of various contaminants, we can efficiently select metal-poor stars and build a very complete sample with high purity. The success rate of uncovering [Fe/H]SEGUE < −3.0 stars among [Fe/H]Pristine < −3.0 selected stars is 24 per cent, and 85 per cent of the remaining candidates are still very metal poor ([Fe/H]<−2.0). We further demonstrate that Pristine is well suited to identify the very rare and pristine Galactic stars with [Fe/H] < −4.0, which can teach us valuable lessons about the early Universe.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,001 | 0,001 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,000 | 0,000 |
| Bibliométrie | 0,003 | 0,003 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,001 | 0,001 |
| Science ouverte | 0,001 | 0,001 |
| Intégrité de la recherche | 0,000 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,002 | 0,004 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».