Detection, communication, and individual identification with deep audio embeddings: A case study with North Atlantic
Irina Tolkova1, Holger Klinck1,2, Dana A Cusano3
1Cornell K. Lisa Yang Center for Conservation Bioacoustics, Cornell University, Ithaca, New York, United States of America.
None:
Anthropogenic noise has increased ambient sound levels across the globe, both underwater and on land. Among its many negative impacts, heightened noise can impair communication in vocal animals through acoustic masking. Conceptually, noise reduces the animal's communication space - the area in which an individual animal can effectively convey information to a conspecific listener. Previous studies have estimated the communication space using sound propagation models and/or behavioral studies. However, studies frequently equate signal recognition with signal detection - a necessary but not sufficient precondition - thereby persistently overestimating spatial coverage and underestimating anthropogenic impacts. Measuring communication is inherently difficult, and varies with taxa, call type, and context, leading to significant data gaps in key parameters. We propose that deep learning creates an opportunity to estimate biologically-relevant communication, even for data-limited species. In particular, we present a case study with the critically endangered North Atlantic right whale (Eubalaena glacialis; hereafter NARW). Prior research has demonstrated that the upcall - a low-frequency contact call produced across ages and sexes - encodes individual identity. We therefore consider a dataset of NARW vocalizations recorded with on-animal archival tags, spanning 234 samples across 11 individuals from 3 sites. First, we demonstrate that audio embeddings from the BirdNET model can robustly distinguish individual right whales. Then, we simulate the effect of varying ambient noise levels to estimate signal excess for both signal detection and individual identification, finding that an additional ≥7 dB is necessary for the model to distinguish individuals. Altogether, we hope this work provides both a methodological advance for individual identification and a framework for better understanding anthropogenic impacts on vocal wildlife.
More Related Videos
Related Concept Videos
Perception of Sound Waves
The pitch of a sound depends on the frequency and the pressure amplitude of the source. Two sounds of the same frequency...
Methods of Classification and Identification
Auditory Perception
Hearing
Detection of Black Holes
Their closest cousins are neutron stars, which are composed almost entirely of neutrons packed against each other, making them extremely dense. A neutron star has the same mass as the Sun but its diameter is only a few kilometers. Therefore, the escape velocity from their surface is close to the speed of light.
Not until the 1960s, when the first neutron...


