Industry Bibliographical Databases: Perspectives of Use in the FMBA of Russia for Scientific Expertise in Decision-Making. Report 1. General Issues and Database on Health and Other Effects in Nuclear Workers
Notice bibliographique
Résumé
The presented review of three reports is devoted to bibliographic databases on health and other effects and indexes in nuclear workers (NW) and uranium miners (U miners), developed within the framework of the research theme of the Federal Medical and Biological Agency of Russia (FMBA) and registered with the state in Rospatent. Report 1 sets out introductory issues of the theory of databases, as well as registers, and provides detailed information on the database for NW. The purpose of the database for NW creating was to form a repository for accessible for abstract and full-text search published data on themes relevant for conducting research examinations for expertise in the system of the FMBA, in other healthcare institutions dealing with the radiation factor, and, more broadly, for conducting fundamental and applied research in the field of professional exposures. The main parts of the database are two separate sub-databases for Russian and foreign NW (Russian NW and Foreign NW), in which the sources are collected in alphabetical order by the authors of the publication or the organizations that created the document. The structural form of information is a catalog that includes primary (main) units of information in the form of an information file about the source (DOC), which contains the title of the publication/document, an abstract (sometimes additional information), and the full original publication (PDF, rarely HTML), available for 88–91 % of sources (the Russian and Foreign sub-databases contain 2078 and 2145 sources, respectively, as of the end of January 2025). 51 % of the works in the database correspond to studies for Russian NW; followed by the USA, Great Britain, Canada, France, and Japan. Visual and/or software search of the material in the database it is supposed to be carried out both through the information names of the catalogs, including the themes of research, carried out using the list of abbreviations (metadata for the database), and through all the texts of the sources included in the database using the proposed programs. Auxiliary elements of the database are fragments of two sub-bases that have undergone hierarchical thematic cataloging in accordance with the identified areas of research on the effects and indexes for NW. These elements are intended, firstly, for initial familiarization with the subject of the database for NW, and, secondly, they are significant as a final thematic base with a certain number of sources, which can be used directly for operational purposes. The developed database for NW has no analogues among industry databases/registers for NW in various countries, nor among bibliographic and search systems. PubMed, Cochrane Library, EMBASE, CINAHL, ISRCTN, Web of Science and Google revealed 5–24 times fewer sources on the theme than the proposed database, and in most cases the world search systems do not provide for the extraction of original publications (as for the IAEA INIS bibliographic database on radiation effects). The depth of the search for works on effects and indexes for NW in the world systems is significantly inferior to the developed database (1960–1970s versus 1940–1950s). It is concluded that the presented database on NW is unique for examination within the framework of the FMBA and other healthcare institutions, and has no complete replacement as a scientific reference and expert depot of sources.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction distillée sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.
Scores Codex et Gemma par catégorie
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,003 | 0,004 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,000 | 0,000 |
| Bibliométrie | 0,001 | 0,001 |
| Études des sciences et des technologies | 0,000 | 0,001 |
| Communication savante | 0,000 | 0,000 |
| Science ouverte | 0,000 | 0,000 |
| Intégrité de la recherche | 0,000 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,000 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».