Bibliographic record
Abstract
1. Summary This data is public for a manuscript under review. 2. File list AlaskaIceResultsv1 (.shp, .dbf, .prj, .shx) River ice breakup, freeze-up, duration, and uncertainty observations for Alaskan rivers wider than 150m from 2000-2018 CanadaValidationResultsv1 (.shp, .dbf, .prj, .shx) River ice breakup, freeze-up, duration, and uncertainty observations for reaches located near Canadian WSC gages from 2000-2018 validationLocationsUSGS (.shp, .dbf, .prj, .shx) Locations of USGS gages used for ice flag validation validationLocationsWSC (.shp, .dbf, .prj, .shx) Locations of WSC gages used for ice flag validation validationLocationsNWS (.shp, .dbf, .prj, .shx) Locations of NWS first ice and breakup observations AlaskanRegions (.shp, .dbf, .prj, .shx) Regions used for regional Alaskan analysis 3. Attribute description AlaskaIceResultsv1 & CanadaValidationResultsv1 id: Alaskan or Canadian reach ID widthMd: Median reach width (m) derived from the GRWL vector product (Allen & Pavelsky, 2018) elevMed (only AlaskaIceResultsv1): Median reach elevation (m) derived from the GRWL vector product (Allen & Pavelsky, 2018) region (only AlaskaIceResultsv1): The region of AK in which the reach resides. iceYear: The ice year of the observation based upon a year starting August 1st and ending July 31st. For example, ice year 2001 refers to Aug 1st 2000 – July 31st 2001. duration: Total ice duration (days). 0 = no duration data rdThresh: Reach-specific red band threshold used to distinguish between ice and water pixels observing the reach fuDate/buDate: Freeze-up/breakup date. Blank = no freeze-up/breakup date fuIcDOY/buIcDOY: Freeze-up/breakup day of the ice year. 0 = no freeze-up/breakup date fuUncrt/buUncrt: Uncertainty in freeze-up/breakup date. 0 = no freeze-up/breakup date fUncertF/bUncertF: Uncertainty flag for freeze-up/breakup dates. “good” = uncertainty <= ±10 days “flag” = uncertainty > ±10 days Blank = no freeze-up/breakup date validationLocationsUSGS & validationLocationsWSC id: Alaskan or Canadian reach ID country: USA or Canada org: United States Geological Survey (USGS) or Water Survey of Canada (WSC) stationNum: Gage station number stationNm: Gage station name yearFrom: Year that discharge data collection begun at the station yearTo: Year that discharge data collection ended at the station validationLocationsNWS id: Alaskan reach ID fnLctn: Name of NWS location nrCnfln: Is the observation near a confluence (y/n) mltChnn: Is the observation on a multichannel river (y/n) nYersFU: Number of years of NWS first-ice dates compared against MODIS freeze-up dates during the validation process nYersBU: Number of years of NWS breakup dates compared against MODIS breakup dates during the validation process AlaskanRegions region: Alaskan region name
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.003 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.003 | 0.005 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.002 | 0.002 |
| Open science | 0.002 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.059 | 0.073 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".