由本地听众对编码的外国口音演讲的口音评分
Jing Yang1, Jaskirat Sidhu1, Gabrielle Totino2
1Communication Sciences and Disorders, University of Wisconsin-Milwaukee, Milwaukee, Wisconsin 53201, USA.
JASA express letters
|September 25, 2023
概括
使用语音编码器进行语音处理,降低了普通话音英语中感知到的外语口音强度. 根据语音编码器频道的数量,口音评分有所不同,对母语和非母语使用者的影响不同.
科学领域:
- 语音学和语音科学 语音学和语音科学
- 声学语音学的声音学
- 获得第二语言的学习.
背景情况:
- 了解外语口音的感知对于语音处理和沟通至关重要.
- 之前的研究已经探索了影响口音判断的因素,但声编码器道操纵效应需要进一步调查.
研究的目的:
- 调查如何操纵语音编码器通道的数量影响普通话音英语语音的感知发音.
- 为了比较听众对语音编码语音的判断与母语和非母语英语使用者的未处理语音.
主要方法:
- 收集了12名普通话口音英语发言人和2名英语母语发言人的语音样本.
- 使用1,2,4,8和16通道的噪音和音调声编码器处理的语音样本.
- 53名英语母语听众在9分级别上对语音编码和未处理的语音进行了评分.
主要成果:
- 与未经处理的情况相比,在语音编码条件下,外国口音的说话者被认为在语音编码条件下口音较弱.
- 随着频道数量的变化,对本地和外国口音发言者观察到不同的口音评级变化模式.
- 减轻口音的程度随着语音编码器频道的数量而变化.
结论:
- 语音编码器处理,特别是使用更少的频道,可以显著降低外国口音的感知强度.
- 声编码器通道缩小对口音感知的影响不均,并且在母语和非母语使用者之间存在差异.
- 这些发现有助于理解外国口音的声学相关性,并为语音合成和修改技术提供信息.
相关概念视频
Air-entraining Agents
92
Air-entraining agents improve the durability and workability of concrete in climates with frequent freezing and thawing. These agents prevent cracks by introducing small air bubbles into the mix, creating spaces accommodating water expansion when temperatures drop. The air-entraining agents lower the surface tension of water, forming stable, small air bubbles. This method is more effective than having accidental large voids, as the intentional, smaller, and evenly distributed air voids improve...
92
Auditory Perception
357
The auditory system is essential for sound perception, utilizing various critical structures. When sound waves enter the outer ear, they travel through the ear canal and cause the eardrum to vibrate. These vibrations are then transmitted to the middle ear, where three tiny bones – the malleus, incus, and stapes – amplify the sound. This amplification is crucial, as it ensures that the sound vibrations are strong enough to be conveyed to the inner ear. These vibrations then reach the...
357
Hearing
52.5K
When we hear a sound, our nervous system is detecting sound waves—pressure waves of mechanical energy traveling through a medium. The frequency of the wave is perceived as pitch, while the amplitude is perceived as loudness.
52.5K
Perceiving Loudness, Pitch, and Location
233
The human brain perceives pitch through two primary mechanisms reflected in place theory and frequency theory. Each mechanism describes how sound waves are interpreted as specific pitches by the brain, offering insights into the intricate processes of auditory perception.
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by...
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by...
233
Sound Intensity Level
4.2K
Humans perceive sound by hearing. The human ear helps sound waves reach the brain, which then interprets the waves and creates the perception of hearing. The loudness of the environment in which a person is located determines whether they can distinguish between different sound sources.
The human ear can perceive an extensive range of sound intensity, necessitating the use of the logarithmic scale to define a physical quantity—the intensity level. It is a ratio of two intensities and...
The human ear can perceive an extensive range of sound intensity, necessitating the use of the logarithmic scale to define a physical quantity—the intensity level. It is a ratio of two intensities and...
4.2K
Components of Language
308
Language, whether spoken, signed, or written, consists of specific components: lexicon and grammar. The lexicon is the vocabulary of a language, comprising its words. Grammar is the set of rules used to convey meaning through the lexicon. For example, English grammar adds “-ed” to most verbs to indicate past tense. Words are formed by combining phonemes, which are the basic sound units of a language. Different languages have different sets of phonemes (e.g., “ah” vs.
308


