Jove
Visualize
Contact Us
JoVE
x logofacebook logolinkedin logoyoutube logo
ABOUT JoVE
OverviewLeadershipBlogJoVE Help Center
AUTHORS
Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
LIBRARIANS
TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
RESEARCH
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
EDUCATION
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
Terms & Conditions of Use
Privacy Policy
Policies

Related Concept Videos

Perceiving Loudness, Pitch, and Location01:21

Perceiving Loudness, Pitch, and Location

1.3K
The human brain perceives pitch through two primary mechanisms reflected in place theory and frequency theory. Each mechanism describes how sound waves are interpreted as specific pitches by the brain, offering insights into the intricate processes of auditory perception.
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by...
1.3K
Auditory Perception01:17

Auditory Perception

1.5K
The auditory system is essential for sound perception, utilizing various critical structures. When sound waves enter the outer ear, they travel through the ear canal and cause the eardrum to vibrate. These vibrations are then transmitted to the middle ear, where three tiny bones – the malleus, incus, and stapes – amplify the sound. This amplification is crucial, as it ensures that the sound vibrations are strong enough to be conveyed to the inner ear. These vibrations then reach the...
1.5K
Perception of Sound Waves01:01

Perception of Sound Waves

6.0K
The human ear is not equally sensitive to all frequencies in the audible range. It may perceive sound waves with the same pressure but different frequencies as having different loudness. Moreover, the perception of sound waves depends on the health of an individual's ears, which decays with age. The health of one's ears may also be affected by regular exposure to loud noises.
The pitch of a sound depends on the frequency and the pressure amplitude of the source. Two sounds of the same...
6.0K
Facial Feedback Hypothesis01:24

Facial Feedback Hypothesis

859
Charles Darwin proposed that facial expressions are an evolutionary adaptation for communication. He argued that these expressions are not influenced by culture but are universal across species. For example, a snarling expression with exposed teeth signals a threat in many animals, including humans. Darwin also suggested that displaying an emotion can intensify the feeling. Smiling, for example, could enhance one's sense of happiness. This idea laid the foundation for understanding the role...
859
Factors Affecting Perception01:25

Factors Affecting Perception

3.1K
Perception is influenced by perceptual set, context, motivation, and emotion. Perceptual set, or perceptual expectancy, refers to the tendency to perceive things in a particular way, influenced by previous experiences and expectations. This phenomenon affects the interpretation of stimuli, creating a set of mental tendencies and assumptions that impact sensory perceptions of sound, taste, touch, and sight.
An illustrative example of a perceptual set is the scenario where an airline pilot told...
3.1K
Self-Evaluation Maintenance Model01:29

Self-Evaluation Maintenance Model

396
The Self-Evaluation Maintenance (SEM) model offers a psychological framework to understand how individuals’ self-esteem is influenced by the achievements of others, particularly those with whom they share close personal bonds. The SEM model operates when personal rather than social identity guides individuals. Central to this model is the notion that individuals have an inherent desire to preserve a favorable self-image, which is continuously shaped by interpersonal comparisons and...
396

You might also read

Related Articles

Articles linked to this work by shared authors, journal, and citation graph.

Sort by
Same author

Improving zero-shot style transfer text-to-speech by disentangled fine-grained style modeling.

JASA express letters·2026
Same author

Introduction to special issue on acoustic cue-based perception and production of speech by humans and machines.

The Journal of the Acoustical Society of America·2025
Same author

Unraveling the associations between voice pitch and major depressive disorder: a multisite genetic study.

Molecular psychiatry·2024
Same author

The JIBO Kids Corpus: A speech dataset of child-robot interactions in a classroom environment.

JASA express letters·2024
Same author

Unraveling the Associations Between Voice Pitch and Major Depressive Disorder: A Multisite Genetic Study.

medRxiv : the preprint server for health sciences·2024
Same author

Speechformer-CTC: Sequential Modeling of Depression Detection with Speech Temporal Classification.

Speech communication·2024

Related Experiment Video

Updated: Apr 6, 2026

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
09:09

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody

Published on: September 27, 2024

980

Perceptual evaluation of voice source models.

Jody Kreiman1, Marc Garellek2, Gang Chen3

  • 1Department of Head and Neck Surgery, University of California-Los Angeles School of Medicine, 31-24 Rehabilitation Center, Los Angeles, California 90095-1794, USA.

The Journal of the Acoustical Society of America
|August 3, 2015
PubMed
Summary

Understanding voice production requires knowing which voice source model features matter to listeners. This study found that how well a model fits a voice pulse doesn't predict perceived voice quality.

More Related Videos

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception
05:48

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception

Published on: August 9, 2024

2.1K
Asthma Detection Research Based on Voice Signal Processing and Machine Learning
04:04

Asthma Detection Research Based on Voice Signal Processing and Machine Learning

Published on: July 22, 2025

1.2K

Related Experiment Videos

Last Updated: Apr 6, 2026

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
09:09

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody

Published on: September 27, 2024

980
Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception
05:48

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception

Published on: August 9, 2024

2.1K
Asthma Detection Research Based on Voice Signal Processing and Machine Learning
04:04

Asthma Detection Research Based on Voice Signal Processing and Machine Learning

Published on: July 22, 2025

1.2K

Area of Science:

  • Phonetics and Speech Science
  • Acoustic Analysis
  • Auditory Perception

Background:

  • Voice source models are crucial for synthesizing realistic speech.
  • Existing models vary in their accuracy (fit) to natural voice production.
  • The perceptual relevance of these model-fit differences remains largely unknown.

Purpose of the Study:

  • To investigate the relationship between the goodness-of-fit of voice source models and the perceived quality of synthesized voices.
  • To determine which aspects of voice source modeling are perceptually salient to listeners.

Main Methods:

  • Five voice source models were fitted to 40 natural voice recordings.
  • Speech stimuli were synthesized using each of the five modeled voice sources.
  • Listeners performed a visual sort-and-rate task to assess perceptual similarity.
  • Multidimensional scaling analysis was used to analyze listener judgments.

Main Results:

  • Neither the overall fit to pulse shape nor landmark points predicted perceived voice quality differences.
  • Models better fit the opening phase of glottal pulses than the closing phase.
  • Perceptual similarity was more strongly predicted by features of the closing phase (negative peak timing/amplitude) than peak glottal opening.

Conclusions:

  • The temporal and amplitude characteristics of the glottal flow derivative's negative peak are important for perceived voice quality.
  • Goodness-of-fit in the time domain for voice source models does not directly translate to perceptual relevance.
  • Further research is needed to identify the specific acoustic features of voice production that listeners prioritize.