Japan Biodiversity Information Initiative (JBIF)'s Efforts to Collect and Publish Biodiversity Information from Japan
Notice bibliographique
Résumé
The Japan Biodiversity Information Initiative (JBIF) was originally established in 2007 as the Global Biodiversity Information Facility (GBIF) Japan National Node to aggregate biodiversity data in Japan and conduct publications through GBIF. JBIF was later renamed after Japan became a GBIF observer, but activities including data publication through GBIF have continued to the present. JBIF operates with the support of the National BioResource Project (NBRP) by the Ministry of Education, Culture, Sports, Science and Technology (MEXT), with collaboration from three institutions: the National Institute of Genetics (NIG), the National Institute for Environmental Studies (NIES), and the National Museum of Nature and Science (NMNS). The NBRP is a project that focuses on the collection, preservation, provision, and enhancement of bioresources. JBIF collects both observation and specimen data and publishes them through GBIF. For domestic data use, a search system for data published by JBIF is available on the JBIF website. Moreover, NMNS managed a museum network called the Science Museum Net (S-Net), and bilingual (Japanese and English) specimen data collected by S-Net is also available via the S-Net portal site. We are working to promote the biodiversity informatics field in Japan through a translation of the GBIF resources, including the website, important documents such as the GBIF Science Review, as well as organize workshops and conferences, primarily targeting students, researchers, museum curators, and local government officials, to facilitate the sharing of information and exchange of opinions on biodiversity information. To date, Japan has published 564 datasets and over 12 million occurrences to GBIF, making it the third-largest contributor of data to GBIF in Asia, following India and Taiwan. Moreover, regarding specimen-based occurrence data, Japan is the largest contributor in Asia. In this presentation, we will introduce JBIF's initiatives and future activities.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,040 | 0,042 |
| Méta-épidémiologie (sens strict) | 0,002 | 0,001 |
| Méta-épidémiologie (sens large) | 0,002 | 0,001 |
| Bibliométrie | 0,021 | 0,021 |
| Études des sciences et des technologies | 0,005 | 0,002 |
| Communication savante | 0,012 | 0,011 |
| Science ouverte | 0,003 | 0,011 |
| Intégrité de la recherche | 0,002 | 0,004 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,014 | 0,011 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».