Characterization of low surface brightness structures in annotated deep images
Bibliographic record
Abstract
Context. The identification and characterization of low surface brightness (LSB) stellar structures around galaxies such as tidal debris of ongoing or past collisions is essential to constrain models of galactic evolution. So far most efforts have focused on the numerical census of samples of varying sizes, either through visual inspection or more recently with deep learning. Detailed analyses including photometry have been carried out for a small number of objects, essentially because of the lack of convenient tools able to precisely characterize tidal structures around large samples of galaxies. Aims. Our goal is to characterize in detail, and in particular obtain quantitative measurements, of LSB structures identified in deep images of samples consisting of hundreds of galaxies. Methods. We developed an online annotation tool that enables contributors to delineate the shapes of diffuse extended stellar structures with precision, as well as artifacts or foreground structures. All parameters are automatically stored in a database which may be queried to retrieve quantitative measurements. We annotated LSB structures around 352 nearby massive galaxies with deep images obtained with the Canada-France-Hawaii Telescope as part of two large programs: Mass Assembly of early-Type GaLAxies with their fine Structures and Ultraviolet Near Infrared Optical Northern Survey/Canada-France Imaging Survey. Each LSB structure was delineated and labeled according to its likely nature: stellar shells, streams associated with a disrupted satellite, tails that formed in major mergers, ghost reflections, or cirrus. Results. From our database containing 8441 annotations, the area, size, median surface brightness, and distance to the host of 228 structures were computed. The results confirm the fact that tidal structures defined as streams are thinner than tails, as expected by numerical simulations. In addition, tidal tails appear to exhibit a higher surface brightness than streams (by about 1 mag), which may be related to different survival times for the two types of collisional debris. We did not detect any tidal feature fainter than 27.5 magarcsec −2 , while the nominal surface brightness limits of our surveys range between 28.3 and 29 magarcsec −2 , a difference that needs to be taken into account when estimating the sensitivity of future surveys to identify LSB structures. Conclusions. We compiled an annotation database of observed LSB structures around nearby massive galaxies including tidal features that may be used for quantitative analysis and as a training set for machine learning algorithms.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.002 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.003 | 0.002 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".