Molecular diversity of dissolved organic matter reflects macroecological patterns in river networks
Bibliographic record
Abstract
Deciphering dissolved organic matter (DOM) molecular complexity is crucial for understanding ecosystem function. Using the continental-scale Worldwide Hydrobiogeochemistry Observation Network for Dynamic Rivers Systems (WHONDRS) Fourier-transform ion cyclotron resonance mass spectrometry (FTICR-MS) dataset, we reveal fundamental scaling patterns of DOM chemodiversity with watershed characteristics. Analysis of 54 river sites shows local and regional watershed features significantly influence DOM chemodiversity (2500–8718 unique formulae), exhibiting consistent scaling patterns across compound classes and a novel latitudinal gradient (decreasing diversity with increasing latitude). Scaling relationships for DOM composition vary by compound class. Crucially, the scaling parameters (B, baseline chemodiversity; Z, sensitivity) are linearly interrelated. This B–Z relationship is most robust for potentially bio-labile carbohydrates (coefficient of determination R 2 ≈ 0.85), diminishing for recalcitrant, plant-derived molecules (such as lignin), and indicates (potential) biolability-dependent coupling between baseline diversity and environmental responsiveness. These quantitative scaling relationships, with scaling exponents ranging from − 2.1 to 2.2 across compound classes, enable prediction of DOM composition across watersheds, offering a framework to understand ecosystem responses to environmental change. This research bridges biogeochemistry and ecology, providing tools to anticipate molecular transformations across scales.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".