Bibliographic record
Abstract
Contributors to this version: Trevor James Smith (@Zeitsperre), Juliette Lavoie (@juliettelavoie), Éric Dupuis (@coxipi), Pascal Bourgault (@aulemahal). Announcements <code>xclim</code> now supports testing against tagged versions of <code>Ouranosinc/xclim-testdata <https://github.com/Ouranosinc/xclim-testdata></code>_ in order to support older versions of <code>xclim</code>. For more information, see the Contributing Guide for more details. (PR/1339). <code>xclim v0.42.0</code> will be the last version to explicitly support Python3.8. (GH/1268, PR/1344). New features and enhancements Two previously private functions for selecting a day of year in a time series when performing calendar conversions are now exposed. (GH/1305, PR/1317). New functions are: <code>xclim.core.calendar.yearly_interpolated_doy</code> <code>xclim.core.calendar.yearly_random_doy</code> <code>scipy</code> is no longer pinned below v1.9 and <code>lmoments3>=1.0.5</code> is now a core dependency and installed by default with <code>pip</code>. (GH/1142, PR/1171). Fix bug on number of bins in <code>xclim.sdba.propeties.spatial_correlogram</code>. (PR/1336) Add <code>resample_before_rl</code> argument to control when resampling happens in <code>maximum_consecutive_{frost|frost_free|dry|tx}_days</code> and in heat indices (in <code>_threshold</code>) (GH/1329, PR/1331) Add <code>xclim.ensembles.make_criteria</code> to help create inputs for the ensemble-reduction methods. (GH/1338, PR/1341). Bug fixes Warnings emitted from regular usage of some indices (<code>snowfall_approximation</code> with <code>method="brown"</code>, <code>effective_growing_degree_days</code>) due to successive <code>convert_units_to</code> calls within their logic have been silenced. (PR/1319). Fixed a bug that prevented the use of the <code>sdba_encode_cf</code> option with xarray 2023.3.0 (PR/1333). Fixed bugs in <code>xclim.core.missing</code> and <code>xclim.sdba.base.Grouper</code> when using pandas 2.0. (PR/1344). Breaking changes The call signatures for <code>xclim.ensembles.create_ensemble</code> and <code>xclim.ensembles._base._ens_align_dataset</code> have been deprecated. Calls to these functions made with the original signature will emit warnings. Changes will become breaking in <code>xclim>=0.43.0</code>.(GH/1305, PR/1317). Affected variable: <code>mf_flag</code> (bool) -> <code>multifile</code> (bool) The indice and indicator for <code>last_spring_frost</code> has been modified to use <code>tasmin</code> by default, reflecting its docstring and literature definition (GH/1324, PR/1325). following indices now accept the <code>op</code> argument for modifying the threshold comparison operator (PR/1325): <code>snw_season_length</code>, <code>snd_season_length</code>, <code>growing_season_length</code>, <code>frost_season_length</code>, <code>frost_free_season_length</code>, <code>rprcptot</code>, <code>daily_pr_intensity</code> In order to support older environments, <code>pandas</code> is now conditionally pinned below v2.0 when installing <code>xclim</code> on systems running Python3.8. (PR/1344). Bug fixes <code>xclim.indices.run_length.last_run</code> nows works when <code>freq</code> is not <code>None</code>. (GH/1321, PR/1323). Internal changes Added <code>xclim</code> to the ouranos Zenodo community . (PR/1313). Significant documentation adjustments. (GH/1305, PR/1308): The CONTRIBUTING page has been moved to the top level of the repository. Information concerning the licensing of xclim is clearly indicated in README. <code>sphinx-autodoc-typehints</code> is now used to simplify call signatures generated in documentation. The SDBA module API is now found with the rest of the User API documentation. <code>HISTORY.rst</code> has been renamed <code>CHANGES.rst</code>, to follow <code>dask</code>-like conventions. Hyperlink targets for individual <code>indices</code> and <code>indicators</code> now point to their entries under <code>API</code> or <code>Indices</code>. Module-level docstrings have migrated from the library scripts directly into the documentation RestructuredText files. The documentation now includes a page explaining the reasons for developing <code>xclim</code> and a section briefly detailing similar and related projects. Markdown explanations in some Jupyter Notebooks have been edited for clarity Removed <code>Mapping</code> abstract base class types in call signatures (<code>dict</code> variables were always expected). (PR/1308). Changes in testing setup now prevent <code>test_mean_radiant_temperature</code> from sometimes causing a segmentation fault. (GH/1303, PR/1315). Addressed a formatting bug that caused <code>Indicators</code> with multiple variables returned to not be properly formatted in the documentation. (GH/1305, PR/1317). <code>tox</code> now include <code>sbck</code> and <code>eofs</code> flags for easier testing of dependencies. CI builds now test against <code>sbck-python</code> @ master. (PR/1328). <code>upstream</code> CI tests are now run on push to master, at midnight, and can also be triggered via <code>workflow_dispatch</code>. Failures from upstream build will open issues using <code>xarray-contrib/issue-from-pytest-log</code>. (PR/1327). Warnings from set <code>_version_deprecated</code> within Indicators now emit <code>FutureWarning</code> instead of <code>DeprecationWarning</code> for greater visibility. (PR/1319). The <code>Graphics</code> section of the <code>Usage</code> notebook has been expanded upon while grammar and spelling mistakes within the notebook-generated documentation have been reduced. (GH/1335, PR/1338, suggested from PyOpenSci Software Review). The Contributing guide now lists three separate subsections to help users understand the gains from optional dependencies. (GH/1335, PR/1338, suggested from PyOpenSci Software Review).
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.002 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".