Related Experiment Video
Updated: May 19, 2026

A Lightweight, Headphones-based System for Manipulating Auditory Feedback in Songbirds
Published on: November 26, 2012
TweetyBERT: Automated parsing of birdsong through self-supervised machine learning
George Vengrovski1,2, Miranda R Hulsey-Vincent1,2, Melissa A Bemrose2
1Institute of Neuroscience and Department of Biology, University of Oregon, Eugene, OR, USA.
None:
Deep neural networks can be trained to parse animal vocalizations-serving to identify the units of communication and annotating sequences of vocalizations for subsequent statistical analysis. However, current methods rely on human-labeled data for training. The challenge of parsing animal vocalizations in a fully unsupervised manner remains an open problem. Addressing this challenge, we introduce TweetyBERT, a self-supervised transformer neural network developed for the analysis of birdsong. The model is trained to predict masked or hidden fragments of audio but is not exposed to human supervision or labels. Applied to canary song, TweetyBERT autonomously learns the behavioral units of song, such as notes, syllables, and phrases-capturing intricate acoustic and temporal patterns. This approach of developing self-supervised models specifically tailored to animal communication may significantly accelerate the analysis of unlabeled vocal data.
More Related Videos
04:04Asthma Detection Research Based on Voice Signal Processing and Machine Learning
Published on: July 22, 2025
06:22Machine Learning-Based Cough Tone Classification: Diagnostic Exploration of Chronic Obstructive Pulmonary Disease and Respiratory Tract Infections
Published on: September 19, 2025
Related Concept Videos
Classification of Signals
A continuous-time signal holds a value at every instant in time, representing information seamlessly. In contrast, a discrete-time signal holds values only at specific moments, often denoted as x(n), where...
Automatic Processing and Automatic Social Behavior