Jove
Visualize
Contact Us
JoVE
x logofacebook logolinkedin logoyoutube logo
ABOUT JoVE
OverviewLeadershipBlogJoVE Help Center
AUTHORS
Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
LIBRARIANS
TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
RESEARCH
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
EDUCATION
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
Terms & Conditions of Use
Privacy Policy
Policies

Related Concept Videos

Higher Mental Functions of the Brain: Language01:10

Higher Mental Functions of the Brain: Language

3.1K
Language is a system of communication that allows the expression of thoughts, ideas, and feelings. The brain processes language in both hemispheres.
Language formation and comprehension take place in the dominant hemisphere. The dominant hemisphere is responsible for understanding the meaning of spoken, written, or sign language, as well as the ability to communicate. For most people, the left hemisphere is the dominant one. The right hemisphere, then, gives tone and emotional context to the...
3.1K
Components of Language01:24

Components of Language

679
Language, whether spoken, signed, or written, consists of specific components: lexicon and grammar. The lexicon is the vocabulary of a language, comprising its words. Grammar is the set of rules used to convey meaning through the lexicon. For example, English grammar adds “-ed” to most verbs to indicate past tense. Words are formed by combining phonemes, which are the basic sound units of a language. Different languages have different sets of phonemes (e.g., “ah” vs.
679
Hearing01:31

Hearing

56.2K
When we hear a sound, our nervous system is detecting sound waves—pressure waves of mechanical energy traveling through a medium. The frequency of the wave is perceived as pitch, while the amplitude is perceived as loudness.
56.2K
Echo01:06

Echo

826
The human ear cannot distinguish between two sources of sound if they happen to reach within a specific time interval, typically 0.1 seconds apart. More than this, and they are perceived as separate sources.
Imagine the sound is reflected back to the ears. Assuming that the source is very close to the human, the difference between hearing the two sounds—the emitted sound and the reflected sound—may be more than the minimum time for perceiving distinct sounds. If this is the case,...
826
Parallel Processing01:20

Parallel Processing

554
The brain processes sensory information rapidly due to parallel processing, which involves sending data across multiple neural pathways at the same time. This method allows the brain to manage various sensory qualities, such as shapes, colors, movements, and locations, all concurrently. For instance, when observing a forest landscape, the brain simultaneously processes the movement of leaves, the shapes of trees, the depth between them, and the various shades of green. This enables a quick and...
554
Auditory Pathway01:15

Auditory Pathway

6.9K
Auditory pathways constitute the complex neural circuits responsible for transmitting and interpreting auditory information from the peripheral auditory system to the brain. Sound waves are initially captured by the outer ear, funneled through the ear canal, and reach the tympanic membrane (eardrum). These vibrations are transmitted via the middle ear's ossicles to the inner ear's cochlea.
When viewed cross-sectionally, the cochlea reveals the scala vestibuli and scala tympani flanking...
6.9K

You might also read

Related Articles

Articles linked to this work by shared authors, journal, and citation graph.

Sort by
Same author

Audiometric detection thresholds for older adults with normal and impaired hearing predict recognition of spectrally and temporally degraded speech in speech-modulated noise.

International journal of audiology·2026
Same author

Exploring Dysphonic Artificial Intelligence Voice Cloning for Speech Intelligibility in Noise.

Journal of voice : official journal of the Voice Foundation·2026
Same author

Parameter efficient speaker adaptation for enhancement of bone-conducted speech.

JASA express letters·2026
Same author

Auditory and cognitive contributions to recognition of degraded speech in noise: Individual differences among older adults.

PloS one·2025
Same author

Release from same-talker speech-in-speech masking: Effects of masker intelligibility and other contributing factorsa).

The Journal of the Acoustical Society of America·2024
Same author

Attenuation and distortion components of age-related hearing loss: Contributions to recognizing temporal-envelope filtered speech in modulated noise.

The Journal of the Acoustical Society of America·2024

Related Experiment Video

Updated: Dec 27, 2025

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
09:09

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody

Published on: September 27, 2024

743

Combining partial information from speech and text.

Daniel Fogerty1, Irraj Iftikhar1, Rachel Madorskiy2

  • 1Department of Communication Sciences and Disorders, University of South Carolina, 1705 College Street, Columbia, South Carolina 29208, USA.

The Journal of the Acoustical Society of America
|March 2, 2020
PubMed
Summary

This study shows that combining partial speech and text information, even with mismatched interruptions, improves sentence recognition. Higher interruption rates led to better performance, especially with multimodal input.

More Related Videos

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception
05:48

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception

Published on: August 9, 2024

1.9K
Eye Tracking During Visually Situated Language Comprehension: Flexibility and Limitations in Uncovering Visual Context Effects
07:36

Eye Tracking During Visually Situated Language Comprehension: Flexibility and Limitations in Uncovering Visual Context Effects

Published on: November 30, 2018

16.3K

Related Experiment Videos

Last Updated: Dec 27, 2025

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
09:09

Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody

Published on: September 27, 2024

743
Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception
05:48

Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception

Published on: August 9, 2024

1.9K
Eye Tracking During Visually Situated Language Comprehension: Flexibility and Limitations in Uncovering Visual Context Effects
07:36

Eye Tracking During Visually Situated Language Comprehension: Flexibility and Limitations in Uncovering Visual Context Effects

Published on: November 30, 2018

16.3K

Area of Science:

  • Auditory perception
  • Speech processing
  • Multimodal integration

Background:

  • Understanding how humans combine sensory information is crucial for speech perception.
  • Partial or degraded auditory information often requires integration with other modalities for comprehension.
  • Interruption rates significantly impact the effectiveness of speech and text information.

Purpose of the Study:

  • To investigate the combination of partial speech and text information at varying interruption rates.
  • To determine how multimodal information supports sentence recognition in quiet conditions.
  • To assess the impact of interruption rate mismatches between modalities.

Main Methods:

  • Speech and text stimuli were presented unimodally or multimodally.
  • Stimuli were interrupted by silence at different rates.
  • Listener performance was evaluated across various interruption rates and conditions.

Main Results:

  • Sentence recognition performance improved with higher interruption rates across all conditions.
  • Multimodal presentations generally provided benefits, even with mismatched interruption rates.
  • Individual differences in unimodal performance influenced the benefit gained from multimodal integration.

Conclusions:

  • Combining partial speech with incomplete visual cues enhances sentence intelligibility.
  • Multimodal integration can compensate for degraded speech in adverse listening conditions.
  • Individual listener capabilities play a role in the effectiveness of multimodal speech enhancement.