Machine learning for underwater laser detection and differentiation of macroalgae and coral
Bibliographic record
Abstract
A better understanding of how spatial distribution patterns in important primary producers and ecosystem service providers such as macroalgae and coral are affected by climate-change and human activity-related events can guide us in anticipating future community and ecosystem response. In-person underwater field surveys are essential in capturing fine and/or subtle details but are rarely simple to orchestrate over large spatial scale (e.g., hundreds of km). In this work, we develop an automated spectral classifier for detection and classification of various macroalgae and coral species through a spectral response dataset acquired in a controlled setting and via an underwater multispectral laser serial imager. Transferable to underwater lidar detection and imaging methods, laser line scanning is known to perform in various types of water in which normal photography and/or video methods may be affected by water optical properties. Using off the shelf components, we show how reflectance and fluorescence responses can be useful in differentiating algal color groups and certain coral genera. Results indicate that while macroalgae show many different genera and species for which differentiation by their spectral response alone would be difficult, it can be reduced to a three color-type/class spectral response problem. Our results suggest that the three algal color groups may be differentiated by their fluorescence response at 580 nm and 685 nm using common 450 nm, 490 nm and 520 nm laser sources, and potentially a subset of these spectral bands would show similar accuracy. There are however classification errors between green and brown types, as they both depend on Chl-a fluorescence response. Comparatively, corals are also very diverse in genera and species, and reveal possible differentiable spectral responses between genera, form (i.e., soft vs. hard), partly related to their emission in the 685 nm range and other shorter wavelengths. Moreover, overlapping substrates and irregular edges are shown to contribute to classification error. As macroalgae are represented worldwide and share similar photopigment assemblages within respective color classes, inter color-class differentiability would apply irrespective of their provenance. The same principle applies to corals, where excitation-emission characteristics should be unchanged from experimental response when investigated in-situ .
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.004 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".