Related Experiment Video
Updated: Jun 11, 2026

Perceptual and Category Processing of the Uncanny Valley Hypothesis' Dimension of Human Likeness: Some Methodological Issues
Published on: June 3, 2013
Categorizing normal and pathological voices: automated and perceptual categorization
Virgilijus Uloza1, Antanas Verikas, Marija Bacauskiene
1Department of Otolaryngology, Kaunas University of Medicine, Kaunas, Lithuania. virgilijus.uloza@kmuk.lt
Objectives:
The aims of the present study were to evaluate the accuracy of an elaborated automated voice categorization system that classified voice signal samples into healthy and pathological classes and to compare it with classification accuracy that was attained by human experts.
Material And Methods:
We investigated the effectiveness of 10 different feature sets in the classification of voice recordings of the sustained phonation of the vowel sound /a/ into the healthy and two pathological voice classes, and proposed a new approach to building a sequential committee of support vector machines (SVMs) for the classification. By applying "genetic search" (a search technique used to find solutions to optimization problems), we determined the optimal values of hyper-parameters of the committee and the feature sets that provided the best performance. Four experienced clinical voice specialists who evaluated the same voice recordings served as experts. The "gold standard" for classification was clinically and histologically proven diagnosis.
Results:
A considerable improvement in the classification accuracy was obtained from the committee when compared with the single feature type-based classifiers. In the experimental investigations that were performed using 444 voice recordings coming from 148 subjects, three recordings from each subject, we obtained the correct classification rate (CCR) of over 92% when classifying into the healthy-pathological voice classes, and over 90% when classifying into three classes (healthy voice and two nodular or diffuse lesion voice classes). The CCR obtained from human experts was about 74% and 60%, respectively.
Conclusion:
When operating under the same experimental conditions, the automated voice discrimination technique based on sequential committee of SVM was considerably more effective than the human experts.
Related Concept Videos
Auditory Perception
Perceiving Loudness, Pitch, and Location
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by identifying...
Automatic Processing and Automatic Social Behavior
Hearing
Perception of Sound Waves
The pitch of a sound depends on the frequency and the pressure amplitude of the source. Two sounds of the same frequency...
Sensory Perception: Organization of the Somatosensory System
The receptor level:
The receptor level is the first stage of sensation. It involves the detection of a stimulus by specialized sensory receptors. The stimulus must arrive within the receptor's receptive field. Next, the receptor converts the energy of the stimulus...
