MétaCan
Menu
Retour à la cohorte
Enregistrement W4393847985 · doi:10.5281/zenodo.5950000

BirdVox-ANAFCC: A dataset for American Northeast Avian Flight Call Classification

2022· dataset· en· W4393847985 sur OpenAlexaboutno aff
Aurora Cramer, Vincent Lostanlen, Bill Evans, Andrew Farnsworth, Justin Salamon, Juan Pablo Bello

Notice bibliographique

RevueZenodo (CERN European Organization for Nuclear Research) · 2022
Typedataset
Langueen
DomaineBiochemistry, Genetics and Molecular Biology
ThématiqueAnimal Vocal Communication and Behavior
Établissements canadiensnon disponible
Organismes subventionnairesnon disponible
Mots-clésGeography

Résumé

récupéré en direct d'OpenAlex

BirdVox-ANAFCC: A dataset for American Northeast Avian Flight Call Classification<br> ===============================================================<br> Version 2.0, February 2022. https://wp.nyu.edu/birdvox <br> Description<br> --------------- BirdVox-ANAFCC is a dataset of short audio waveforms, each of them containing a flight call from one of 14 birds of North America: four American sparrows, one cardinal, two thrushes, and seven New World warblers.<br> * American Tree Sparrow (ATSP)<br> * Chipping Sparrow (CHSP)<br> * Savannah Sparrow (SAVS)<br> * White-throated Sparrow (WTSP)<br> * Red-breasted Grosbeak (RBGR)<br> * Gray-cheeked Thrush (GCTH)<br> * Swainson's Thrush (SWTH)<br> * American Redstart (AMRE)<br> * Bay-breasted Warbler (BBWA)<br> * Black-throated Blue Warbler (BTBW)<br> * Canada Warbler (CAWA)<br> * Common Yellowthroat (COYE)<br> * Mourning Warbler (MOWA)<br> * Ovenbird (OVEN) It also contains other sounds which are often confused for one of the species above. These "confounding factors" encompass flight calls from other species of birds, vocalizations from non-avian animals, as well as some machine beeps. BirdVox-ANAFCC results from an aggregation of various smaller datasets, integrated under a common taxonomy. For more details on this taxonomy, we refer the reader to [1]: [1] Cramer, Lostanlen, Salamon, Farnsworth, Bello. Chirping up the right tree: Incorporating biological taxonomies into deep bioacoustic classifiers. Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), 2020. The second version of the BirdVox-ANAFCC dataset (v2.0) contains flight calls from the BirdVox-full-night dataset. These flight calls were present in the ICASSP 2020 benchmark but did not appear in the initial release of BirdVox-ANAFCC. <br> Data Files<br> ------------<br> BirdVox-ANAFCC contains the recordings as HDF5 files, sampled at 22,050 Hz, with a single channel (mono). Each HDF5 file contains flight call vocalizations of a particular species. The name of each HDF5 file follows the format: `&lt;data-source&gt;_&lt;taxonomy-code&gt;_original.h5`. The name of the HDF5 dataset in each file is "waveforms", with the corresponding key for each audio recording varying in format depending on the data source. Metadata Files<br> ---------------<br> `taxonomy.yaml` details the three-level taxonomy structure used in this dataset, reflected in three-number-codes which largely follow "&lt;family&gt;.&lt;order&gt;.&lt;species&gt;". Additionally, at any level of the taxonomy, the numeric code "0" is reserved for "other" and the code "X" refers to unknown. For example, 1.1.0 corresponds to an American Sparrow with a species outside of our scope of interest, and 1.1.X corresponds to an American Sparrow of unknown species. At the top level (family), the "other" codes (0.\*.\*) deviate from the family-order-species in order to capture a variety of other out-of-scope sounds, including anthropophony, non-avian biophony, and biophony of avians outside of the scope of interest. <br> Please acknowledge BirdVox-ANAFCC in academic research<br> -------------------------------------------------------------------------- When BirdVox-ANAFCC is used for academic research, we would highly appreciate it if scientific publications of works partly based on this dataset cite the following publication: Cramer, Lostanlen, Salamon, Farnsworth, Bello. Chirping up the right tree: Incorporating biological taxonomies into deep bioacoustic classifiers. Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), 2020. The creation of this dataset was supported by NSF grants 1125098 (BIRDCAST) and 1633259 (BIRDVOX), a Google Faculty Award, the Leon Levy Foundation, and two anonymous donors. Conditions of Use<br> ---------------------- Dataset created by Aurora Cramer, Vincent Lostanlen, Bill Evans, Andrew Farnsworth, Justin Salamon, and Juan Pablo Bello.<br> <br> The BirdVox-ANAFCC dataset is offered free of charge under the terms of the Creative Commons Attribution International License:<br> https://creativecommons.org/licenses/by/4.0/<br> <br> The dataset and its contents are made available on an "as is" basis and without warranties of any kind, including without limitation satisfactory quality and conformity, merchantability, fitness for a particular purpose, accuracy or completeness, or absence of errors. Subject to any liability that may not be excluded or limited by law, the authors are not liable for, and expressly exclude all liability for, loss or damage however and whenever caused to anyone by any use of the BirdVox-ANAFCC dataset or any part of it. <br> Feedback<br> ------------- Please help us improve BirdVox-full-night by sending your feedback to:<br> vincent.lostanlen@gmail.com and auroracramer@nyu.edu In case of a problem, please include as many details as possible.<br> <br> <br> Versions<br> ------------<br> 1.0, May 2020: initial version, paired with ICASSP 2020 publication.<br> 2.0, February 2022: added a missing dataset file (BirdVox-70k), updated name of first author (Aurora Cramer).<br> <br> Acknowledgement<br> --------------------------<br> Jessie Barry, Ian Davies, Tom Fredericks, Jeff Gerbracht, Sara Keen, Holger Klinck, Anne Klingensmith, Ray Mack, Peter Marchetto, Ed Moore, Matt Robbins, Ken Rosenberg, and Chris Tessaglia-Hymes. We thank contributors and maintainers of the Macaulay Library and the Xeno-Canto website. We acknowledge that the land on which the data was collected is the unceded territory of the Cayuga nation, which is part of the Haudenosaunee (Iroquois) confederacy.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction distillée sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.

score de la tête « metaresearch » (Codex)0,001
score de la tête « metaresearch » (Gemma)0,000
Version: codex-gemma-dda1882f352aStatut de validation: machine_predicted_unvalidated
Catégories candidatesMéta-épidémiologie (sens strict), Études des sciences et des technologies, Charge utile insuffisante (le modèle a refusé de juger)
Catégories consensuellesCharge utile insuffisante (le modèle a refusé de juger)
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Sans objet · Signal consensuel: Sans objet
GenreSignal candidat: Jeu de données · Signal consensuel: Jeu de données
Score de désaccord entre enseignants0,016
Score d'incertitude au seuil1,000

Scores Codex et Gemma par catégorie

CatégorieCodexGemma
Métarecherche0,0010,000
Méta-épidémiologie (sens strict)0,0000,000
Méta-épidémiologie (sens large)0,0000,000
Bibliométrie0,0000,000
Études des sciences et des technologies0,0020,000
Communication savante0,0000,000
Science ouverte0,0020,002
Intégrité de la recherche0,0000,000
Charge utile insuffisante (le modèle a refusé de juger)0,0170,001

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,055
Tête enseignante GPT0,303
Écart entre enseignants0,248 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; les deux têtes enseignantes s’accordent sur ce qui est montré ici.

Devis d'étudeSans objet
Domainenon disponible
GenreJeu de données

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations1
Publié2022
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueZenodo (CERN European Organization for Nuclear Research)Même sujetAnimal Vocal Communication and BehaviorTravaux en français237 207