MétaCan
Menu
Back to cohort
Record W4393847985 · doi:10.5281/zenodo.5950000

BirdVox-ANAFCC: A dataset for American Northeast Avian Flight Call Classification

2022· dataset· en· W4393847985 on OpenAlexaboutno aff
Aurora Cramer, Vincent Lostanlen, Bill Evans, Andrew Farnsworth, Justin Salamon, Juan Pablo Bello

Bibliographic record

VenueZenodo (CERN European Organization for Nuclear Research) · 2022
Typedataset
Languageen
FieldBiochemistry, Genetics and Molecular Biology
TopicAnimal Vocal Communication and Behavior
Canadian institutionsnot available
Fundersnot available
KeywordsGeography

Abstract

fetched live from OpenAlex

BirdVox-ANAFCC: A dataset for American Northeast Avian Flight Call Classification<br> ===============================================================<br> Version 2.0, February 2022. https://wp.nyu.edu/birdvox <br> Description<br> --------------- BirdVox-ANAFCC is a dataset of short audio waveforms, each of them containing a flight call from one of 14 birds of North America: four American sparrows, one cardinal, two thrushes, and seven New World warblers.<br> * American Tree Sparrow (ATSP)<br> * Chipping Sparrow (CHSP)<br> * Savannah Sparrow (SAVS)<br> * White-throated Sparrow (WTSP)<br> * Red-breasted Grosbeak (RBGR)<br> * Gray-cheeked Thrush (GCTH)<br> * Swainson's Thrush (SWTH)<br> * American Redstart (AMRE)<br> * Bay-breasted Warbler (BBWA)<br> * Black-throated Blue Warbler (BTBW)<br> * Canada Warbler (CAWA)<br> * Common Yellowthroat (COYE)<br> * Mourning Warbler (MOWA)<br> * Ovenbird (OVEN) It also contains other sounds which are often confused for one of the species above. These "confounding factors" encompass flight calls from other species of birds, vocalizations from non-avian animals, as well as some machine beeps. BirdVox-ANAFCC results from an aggregation of various smaller datasets, integrated under a common taxonomy. For more details on this taxonomy, we refer the reader to [1]: [1] Cramer, Lostanlen, Salamon, Farnsworth, Bello. Chirping up the right tree: Incorporating biological taxonomies into deep bioacoustic classifiers. Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), 2020. The second version of the BirdVox-ANAFCC dataset (v2.0) contains flight calls from the BirdVox-full-night dataset. These flight calls were present in the ICASSP 2020 benchmark but did not appear in the initial release of BirdVox-ANAFCC. <br> Data Files<br> ------------<br> BirdVox-ANAFCC contains the recordings as HDF5 files, sampled at 22,050 Hz, with a single channel (mono). Each HDF5 file contains flight call vocalizations of a particular species. The name of each HDF5 file follows the format: `&lt;data-source&gt;_&lt;taxonomy-code&gt;_original.h5`. The name of the HDF5 dataset in each file is "waveforms", with the corresponding key for each audio recording varying in format depending on the data source. Metadata Files<br> ---------------<br> `taxonomy.yaml` details the three-level taxonomy structure used in this dataset, reflected in three-number-codes which largely follow "&lt;family&gt;.&lt;order&gt;.&lt;species&gt;". Additionally, at any level of the taxonomy, the numeric code "0" is reserved for "other" and the code "X" refers to unknown. For example, 1.1.0 corresponds to an American Sparrow with a species outside of our scope of interest, and 1.1.X corresponds to an American Sparrow of unknown species. At the top level (family), the "other" codes (0.\*.\*) deviate from the family-order-species in order to capture a variety of other out-of-scope sounds, including anthropophony, non-avian biophony, and biophony of avians outside of the scope of interest. <br> Please acknowledge BirdVox-ANAFCC in academic research<br> -------------------------------------------------------------------------- When BirdVox-ANAFCC is used for academic research, we would highly appreciate it if scientific publications of works partly based on this dataset cite the following publication: Cramer, Lostanlen, Salamon, Farnsworth, Bello. Chirping up the right tree: Incorporating biological taxonomies into deep bioacoustic classifiers. Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), 2020. The creation of this dataset was supported by NSF grants 1125098 (BIRDCAST) and 1633259 (BIRDVOX), a Google Faculty Award, the Leon Levy Foundation, and two anonymous donors. Conditions of Use<br> ---------------------- Dataset created by Aurora Cramer, Vincent Lostanlen, Bill Evans, Andrew Farnsworth, Justin Salamon, and Juan Pablo Bello.<br> <br> The BirdVox-ANAFCC dataset is offered free of charge under the terms of the Creative Commons Attribution International License:<br> https://creativecommons.org/licenses/by/4.0/<br> <br> The dataset and its contents are made available on an "as is" basis and without warranties of any kind, including without limitation satisfactory quality and conformity, merchantability, fitness for a particular purpose, accuracy or completeness, or absence of errors. Subject to any liability that may not be excluded or limited by law, the authors are not liable for, and expressly exclude all liability for, loss or damage however and whenever caused to anyone by any use of the BirdVox-ANAFCC dataset or any part of it. <br> Feedback<br> ------------- Please help us improve BirdVox-full-night by sending your feedback to:<br> vincent.lostanlen@gmail.com and auroracramer@nyu.edu In case of a problem, please include as many details as possible.<br> <br> <br> Versions<br> ------------<br> 1.0, May 2020: initial version, paired with ICASSP 2020 publication.<br> 2.0, February 2022: added a missing dataset file (BirdVox-70k), updated name of first author (Aurora Cramer).<br> <br> Acknowledgement<br> --------------------------<br> Jessie Barry, Ian Davies, Tom Fredericks, Jeff Gerbracht, Sara Keen, Holger Klinck, Anne Klingensmith, Ray Mack, Peter Marchetto, Ed Moore, Matt Robbins, Ken Rosenberg, and Chris Tessaglia-Hymes. We thank contributors and maintainers of the Macaulay Library and the Xeno-Canto website. We acknowledge that the land on which the data was collected is the unceded territory of the Cayuga nation, which is part of the Haudenosaunee (Iroquois) confederacy.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.001
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMeta-epidemiology (narrow), Science and technology studies, Insufficient payload (model declined to judge)
Consensus categoriesInsufficient payload (model declined to judge)
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: Not applicable
GenreCandidate signal: Dataset · Consensus signal: Dataset
Teacher disagreement score0.016
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0010.000
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0000.000
Science and technology studies0.0020.000
Scholarly communication0.0000.000
Open science0.0020.002
Research integrity0.0000.000
Insufficient payload (model declined to judge)0.0170.001

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.055
GPT teacher head0.303
Teacher spread0.248 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; both teacher heads agree on what is shown here.

Study designNot applicable
Domainnot available
GenreDataset

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations1
Published2022
Admission routes1
Has abstractyes

Explore more

Same venueZenodo (CERN European Organization for Nuclear Research)Same topicAnimal Vocal Communication and BehaviorFrench-language works237,207