MétaCan
Menu
Retour à la cohorte
Enregistrement W7135417822 · doi:10.5523/bris.3g8dm9c6z4tfm2ht6c7l0t43ul

The PanAf-FGBG Dataset

2025· dataset· W7135417822 sur OpenAlexaboutno aff
Otto Brookes, Maksim Kukushkin, Majid Mirmehdi, Colleen Stephens, Paula Dieguez, Thurston C. Hicks, Sorrel Jones, Kevin Lee, Maureen S. McCarthy, Amelia Meier, Emmanuelle Normand, Erin G. Wessling, Roman M.Wittig, Kevin E. Langergraber, Klaus Zuberbuhler, Lukas Boesch, Thomas Schmid, Mimi Arandjelovic, Hjalmar S. Kühl, Tilo Burghardt

Notice bibliographique

RevueBristol Research (University of Bristol) · 2025
Typedataset
Langue
Domaine
Thématique
Établissements canadiensnon disponible
Organismes subventionnairesnon disponible
Mots-clésMetadataCitationEndangered speciesWildlifeResource (disambiguation)Geocoding

Résumé

récupéré en direct d'OpenAlex

DESCRIPTION. The PanAf-FGBG dataset comprises behaviour-annotated video footage of wild chimpanzees from more than 350 camera locations across tropical Africa, collected by the Pan African Programme: The Cultured Chimpanzee. It includes paired foreground (with chimpanzees) and background (without chimpanzees) videos, allowing controlled analysis of background influence on behaviour recognition models. The dataset is split into overlapping and disjoint camera location views to support evaluation under both in-distribution and out-of-distribution conditions. Each entry is accompanied by metadata and multi-label annotations for 14 distinct behaviours, enabling robust model training and testing. This resource aims to enhance AI models for wildlife behaviour understanding and supports broader conservation efforts for endangered great ape species. CITATION. When using this data please cite this dataset deposit and the associated paper where the dataset and baselines are explained in detail: "The PanAf-FGBG Dataset: Understanding the Impact of Backgrounds in Wildlife Behaviour Recognition" published in the 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) available here: https://openaccess.thecvf.com/content/CVPR2025/papers/Brookes_The_PanAf-FGBG_Dataset_Understanding_the_Impact_of_Backgrounds_in_Wildlife_CVPR_2025_paper.pdf. For BIBTEX citation details please see the project website: https://obrookes.github.io/panaf-fgbg.github.io/ ACKNOWLEDGEMENTS. We thank the Pan African Programme: 'The Cultured Chimpanzee' team and its collaborators for allowing the use of their data for this paper. We thank Amelie Pettrich, Antonio Buzharevski, Eva Martinez Garcia, Ivana Kirchmair, Sebastian Schütte, Linda Gerlach and Fabina Haas. We also thank management and support staff across all sites; specifically Yasmin Moebius, Geoffrey Muhanguzi, Martha Robbins, Henk Eshuis, Sergio Marrocoli and John Hart. Thanks to the team at https://www.chimpandsee.org particularly Briana Harder, Anja Landsmann, Laura K. Lynn, Zuzana Macháčková, Heidi Pfund, Kristeena Sigler and Jane Widness. The work that allowed for the collection of the dataset was funded by the Max Planck Society, Max Planck Society Innovation Fund, and Heinz L. Krekeler. In this respect we would like to thank: Ministre des Eaux et Forêts, Ministère de l'Enseignement supérieur et de la Recherche scientifique in Côte d'Ivoire; Institut Congolais pour la Conservation de la Nature, Ministère de la Recherche Scientifique in Democratic Republic of Congo; Forestry Development Authority in Liberia; Direction Des Eaux Et Forêts, Chasses Et Conservation Des Sols in Senegal; Makerere University Biological Field Station, Uganda National Council for Science and Technology, Uganda Wildlife Authority, National Forestry Authority in Uganda; National Institute for Forestry Development and Protected Area Management, Ministry of Agriculture and Forests, Ministry of Fisheries and Environment in Equatorial Guinea. This work was supported by the UKRI CDT in Interactive AI (grant EP/S022937/1). This work was in part supported by the US National Science Foundation Awards No. 2118240 "HDR Institute: Imageomics: A New Frontier of Biological Information Powered by Knowledge-Guided Machine Learning" and Award No. 2330423 and Natural Sciences and Engineering Research Council of Canada under Award No. 585136 for the "AI and Biodiversity Change (ABC) Global Center". WEBSITE. Further materials are available at the project website at: https://obrookes.github.io/panaf-fgbg.github.io/

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,001
score de la tête « metaresearch » (Gemma)0,005
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Sans objet · Signal consensuel: Sans objet
GenreSignal candidat: Jeu de données · Signal consensuel: Jeu de données
Score de désaccord entre enseignants0,051
Score d'incertitude au seuil0,103

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,0010,005
Méta-épidémiologie (sens strict)0,0040,001
Méta-épidémiologie (sens large)0,0020,002
Bibliométrie0,0040,005
Études des sciences et des technologies0,0020,001
Communication savante0,0030,003
Science ouverte0,0040,002
Intégrité de la recherche0,0040,002
Charge utile insuffisante (le modèle a refusé de juger)0,0310,060

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,075
Tête enseignante GPT0,370
Écart entre enseignants0,295 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeSans objet
Domainenon disponible
GenreJeu de données

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations0
Publié2025
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueBristol Research (University of Bristol)Travaux en français237 207