Systematically Measuring Ultra-diffuse Galaxies (SMUDGes). I. Survey Description and First Results in the Coma Galaxy Cluster and Environs
Bibliographic record
Abstract
We present a homogeneous catalog of 275 large (effective radius ≳5.″3) ultra-diffuse galaxy (UDG) candidates lying within an ≈290 square degree region surrounding the Coma Cluster. The catalog results from our automated postprocessing of data from the Legacy Surveys, a three-band imaging survey covering 14,000 square degrees of the extragalactic sky. We describe a pipeline that identifies UDGs and provides their basic parameters. The survey is as complete in these large UDGs as previously published UDG surveys of the central region of the Coma Cluster. We conclude that the majority of our detections are at roughly the distance of the Coma Cluster, implying effective radii ≥2.5 kpc, and that our sample contains a significant number of analogs of DF44, where the effective radius exceeds 4 kpc, both within the cluster and in the surrounding field. The g - z color of our UDGs spans a large range, suggesting that even large UDGs may reflect a range of formation histories. A majority of the UDGs are consistent with being lower stellar mass analogs of red sequence galaxies, but we find both red and blue UDG candidates in the vicinity of the Coma Cluster and a relative overabundance of blue UDG candidates in the lower-density environments and the field. Our eventual processing of the full Legacy Surveys data will produce the largest, most homogeneous sample of large UDGs.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".