Tunable Q-factor Wavelet Transform-Based Features in the Classification of Phonation Types in the Singing and

Kiran Reddy Mittapalle1, Paavo Alku1

  • 1Department of Information and Communications Engineering, Aalto University, Espoo 02150, Finland.

Summary

This study introduces a new method using tunable Q-factor wavelet transform (TQWT) to classify voice phonation types. The TQWT-based features significantly improved the accuracy of distinguishing between breathy, neutral, and pressed phonation in both singing and speaking voices.

Related Concept Videos

Wave Parameters01:10

Wave Parameters

The simplest mechanical waves are associated with simple harmonic motion and repeat themselves for several cycles. These simple harmonic waves can be modeled using a combination of sine and cosine functions. Consider a simplified surface water wave that moves across the water's surface. Unlike complex ocean waves, in surface water waves, water moves vertically, oscillating up and down, whereas the disturbance of the wave moves horizontally through the medium. If a seagull is floating on the...
7.6K
Design Example01:23

Design Example

The innovation of touch-tone telephony revolutionized the telecommunications industry by replacing the traditional rotary dial with a dual-tone multi-frequency (DTMF) signaling system. This system uses a matrix-style keypad with buttons arranged in four rows and three columns, creating 12 distinct signals each assigned to a pair of frequencies. Each button press results in a simultaneous generation of two sinusoidal tones – one from a low-frequency group (697 to 941 Hz) and one from a...
316
Perception of Sound Waves01:01

Perception of Sound Waves

The human ear is not equally sensitive to all frequencies in the audible range. It may perceive sound waves with the same pressure but different frequencies as having different loudness. Moreover, the perception of sound waves depends on the health of an individual's ears, which decays with age. The health of one's ears may also be affected by regular exposure to loud noises.
The pitch of a sound depends on the frequency and the pressure amplitude of the source. Two sounds of the same...
4.4K
Sound Waves: Resonance01:14

Sound Waves: Resonance

Resonance is produced depending on the boundary conditions imposed on a wave. Resonance can be produced in a string under tension with symmetrical boundary conditions (i.e., has a node at each end). A node is defined as a fixed point where the string does not move. The symmetrical boundary conditions result in some frequencies resonating and producing standing waves, while other frequencies interfere destructively. Sound waves can resonate in a hollow tube, and the frequencies of the sound...
2.5K
Larynx01:21

Larynx

The human larynx, often referred to as the voice box, is an intricate organ located in the neck. It serves as a pathway for air to enter the lungs during respiration and is an essential component of voice production.
Anatomy of the Larynx
The larynx consists of various components, including cartilage, muscles, and vocal cords. Its structure includes three large unpaired cartilages—the thyroid, cricoid, and epiglottis—and three smaller paired cartilages—the arytenoids,...
1.2K
IR Spectrum Peak Splitting: Symmetric vs Asymmetric Vibrations01:08

IR Spectrum Peak Splitting: Symmetric vs Asymmetric Vibrations

Identical bonds within a polyatomic group can stretch symmetrically (in-phase) or asymmetrically (out-of-phase). Similar to hydrogen bonding, these vibrations also influence the shape of the IR peak. Generally, asymmetric stretching frequencies are higher than symmetric stretching frequencies. For example, primary amines exhibit two distinct IR peaks between 3300–3500 cm−1 corresponding to the symmetric and asymmetric N-H stretching, while secondary amines exhibit a single...
913