The reliability of acoustic classification models for determining avian vocalisation patterns
Metcalf, O. C.; Alencar Nunes, C.; Hopping, W. A.; Lees, A. C.; Lostanlen, V.; Barlow, J.
Show abstract
Automated detection and classification of species vocalisations offers the potential to utilise acoustic datasets across unprecedented spatial and temporal scales. However, classification algorithms inevitably generate errors, and error rates vary with context. While methods for quantifying error rates in ecoacoustics are well established, there is limited research on what level of model performance is sufficient to reliably determine avian vocalisation patterns. Using an extensive fully expert-labelled acoustic dataset from Peru (18 hours, 6 sites), we examined changes in the probability of detecting target bird vocalisations in the first hour after dawn to address three key questions: (1) How sensitive are models predicting detection probability over time to reductions in classification accuracy? (2) To what extent does aggregating detections over longer time periods impact classification accuracy? Additionally, we used the labelled dataset to assess how the creation and composition of a test dataset for assessing classifier performance can impact the reliability of accuracy metrics: (3) Are estimates of classification precision robust when test and deployment datasets are not independent and identically distributed? Our results indicate that poor classification performance--especially low precision--can lead to misleading inferences about temporal patterns of vocalisations. Aggregating classifier predictions over longer time periods improved recall but often resulted in misleading patterns of vocal behaviour by reducing precision and temporal resolution. We also demonstrate that precision can be substantially overestimated when species presences are rarer in the deployment dataset than in test data. These findings highlight the importance of cautious application of automated classification in acoustic ecology and the need for accuracy assessment methods tailored to the intended ecological analysis.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Automated classification of bat echolocation call recordings with artificial intelligence 95%
- Using Neural Style Transfer to study the evolution of animal signal design: A case study in an ornamented fish 92%
- Using machine learning to count Antarctic shag (Leucocarbo bransfieldensis) nests on images captured by Remotely Piloted Aircraft Systems 91%
Similar papers in this journal
- A model-based hypothesis framework to define and estimate the diel niche via the 'Diel.Niche' R package 93%
- Songbird parents coordinate offspring provisioning at fine spatio-temporal scales. 92%
- Roadkill islands: carnivore extinction shifts seasonal use of roadside carrion by generalist avian scavenger 91%
Similar papers in this journal
- callsync: an R package for alignment and analysis of multi-microphone animal recordings 92%
- Both environmental conditions and intra- and interspecific interactions influence the movements of a marine predator 92%
- A user-friendly guide to using distance measures to compare time series in ecology. 92%
Similar papers in this journal
- Testing the maintenance of natural responses to survival-relevant calls in the conservation breeding population of a critically endangered corvid (Corvus hawaiiensis) 94%
- Food-associated calls in disc-winged bats 93%
- Allopatric montane wren-babblers exhibit similar song notes but divergent vocal sequences 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.