• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

重新合成的/hVd/语音的识别:共振峰轮廓的影响

Identification of resynthesized /hVd/ utterances: effects of formant contour.

作者信息

Hillenbrand J M, Nearey T M

机构信息

Department of Speech Pathology and Audiology, Western Michigan University, Kalamazoo 49008, USA.

出版信息

J Acoust Soc Am. 1999 Jun;105(6):3509-23. doi: 10.1121/1.424676.

DOI:10.1121/1.424676
PMID:10380673
Abstract

The purpose of this study was to examine the role of formant frequency movements in vowel recognition. Measurements of vowel duration, fundamental frequency, and formant contours were taken from a database of acoustic measurements of 1668 /hVd/ utterances spoken by 45 men, 48 women, and 46 children [Hillenbrand et al., J. Acoust. Soc. Am. 97, 3099-3111 (1995)]. A 300-utterance subset was selected from this database, representing equal numbers of 12 vowels and approximately equal numbers of tokens produced by men, women, and children. Listeners were asked to identify the original, naturally produced signals and two formant-synthesized versions. One set of "original formant" (OF) synthetic signals was generated using the measured formant contours, and a second set of "flat formant" (FF) signals was synthesized with formant frequencies fixed at the values measured at the steadiest portion of the vowel. Results included: (a) the OF synthetic signals were identified with substantially greater accuracy than the FF signals; and (b) the naturally produced signals were identified with greater accuracy than the OF synthetic signals. Pattern recognition results showed that a simple approach to vowel specification based on duration, steady-state F0, and formant frequency measurements at 20% and 80% of vowel duration accounts for much but by no means all of the variation in listeners' labeling of the three types of stimuli.

摘要

本研究的目的是考察共振峰频率变化在元音识别中的作用。元音时长、基频和共振峰轮廓的测量数据取自一个声学测量数据库,该数据库包含45名男性、48名女性和46名儿童说出的1668个/hVd/话语 [希伦布兰德等人,《美国声学学会杂志》97, 3099 - 3111 (1995)]。从该数据库中选取了一个包含300个话语的子集,其中12个元音数量相等,并且男性、女性和儿童说出的话语样本数量大致相等。要求听众识别原始的自然产生的信号以及两个共振峰合成版本。一组“原始共振峰”(OF)合成信号是使用测量得到的共振峰轮廓生成的,另一组“平坦共振峰”(FF)信号是通过将共振峰频率固定在元音最稳定部分测量得到的值来合成的。结果包括:(a) OF合成信号的识别准确率显著高于FF信号;(b) 自然产生的信号的识别准确率高于OF合成信号。模式识别结果表明,一种基于时长、稳态F0以及元音时长20%和80%处的共振峰频率测量的简单元音指定方法,能够解释听众对这三种类型刺激的标注中大部分但绝非全部的变化。

相似文献

1
Identification of resynthesized /hVd/ utterances: effects of formant contour.重新合成的/hVd/语音的识别:共振峰轮廓的影响
J Acoust Soc Am. 1999 Jun;105(6):3509-23. doi: 10.1121/1.424676.
2
Some effects of duration on vowel recognition.时长对元音识别的一些影响。
J Acoust Soc Am. 2000 Dec;108(6):3013-22. doi: 10.1121/1.1323463.
3
Effects of consonant environment on vowel formant patterns.辅音环境对元音共振峰模式的影响。
J Acoust Soc Am. 2001 Feb;109(2):748-63. doi: 10.1121/1.1337959.
4
Identification of steady-state vowels synthesized from the Peterson and Barney measurements.从彼得森和巴尼的测量数据中合成的稳态元音的识别。
J Acoust Soc Am. 1993 Aug;94(2 Pt 1):668-74. doi: 10.1121/1.406884.
5
The relative importance of spectral tilt in monophthongs and diphthongs.单元音和双元音中频谱倾斜的相对重要性。
J Acoust Soc Am. 2005 Mar;117(3 Pt 1):1395-404. doi: 10.1121/1.1861158.
6
Thresholds for second formant transitions in front vowels.前元音中第二共振峰过渡的阈值。
J Acoust Soc Am. 2005 Nov;118(5):3252-60. doi: 10.1121/1.2074667.
7
Perceptual separation of simultaneous vowels: within and across-formant grouping by F0.同时发出的元音的感知分离:通过基频进行共振峰内部和跨共振峰分组
J Acoust Soc Am. 1993 Jun;93(6):3454-67. doi: 10.1121/1.405675.
8
Vowel formant discrimination for high-fidelity speech.高保真语音的元音共振峰辨别
J Acoust Soc Am. 2004 Aug;116(2):1224-33. doi: 10.1121/1.1768958.
9
The relative contributions of speaking fundamental frequency and formant frequencies to gender identification based on isolated vowels.基于孤立元音的基频和共振峰频率在性别识别中的相对贡献。
J Voice. 2005 Dec;19(4):544-54. doi: 10.1016/j.jvoice.2004.10.006.
10
Synthesis fidelity and time-varying spectral change in vowels.元音的合成保真度与时变频谱变化
J Acoust Soc Am. 2005 Feb;117(2):886-95. doi: 10.1121/1.1852549.

引用本文的文献

1
Speech sound discrimination in background noise across the lifespan: a comparative study in Mongolian gerbils and humans.不同年龄段在背景噪声中的语音辨别能力:蒙古沙鼠与人类的比较研究
Front Aging Neurosci. 2025 Jun 9;17:1570305. doi: 10.3389/fnagi.2025.1570305. eCollection 2025.
2
Evaluating normalization accounts against the dense vowel space of Central Swedish.根据瑞典中部密集元音空间评估归一化账户。
Front Psychol. 2023 Jun 21;14:1165742. doi: 10.3389/fpsyg.2023.1165742. eCollection 2023.
3
Spectrally specific temporal analyses of spike-train responses to complex sounds: A unifying framework.
对复杂声音的尖峰序列反应进行频谱特异性时间分析:一个统一框架。
PLoS Comput Biol. 2021 Feb 22;17(2):e1008155. doi: 10.1371/journal.pcbi.1008155. eCollection 2021 Feb.
4
Nonlinear auditory models yield new insights into representations of vowels.非线性听觉模型为元音表征带来了新的见解。
Atten Percept Psychophys. 2019 May;81(4):1034-1046. doi: 10.3758/s13414-018-01644-w.
5
Token Frequency Effects in Homophone Production: An Elicitation Study.同音异形词生成中的词频效应:一项诱发研究。
Lang Speech. 2018 Sep;61(3):466-479. doi: 10.1177/0023830917737108. Epub 2017 Nov 3.
6
Acoustic Properties Predict Perception of Unfamiliar Dutch Vowels by Adult Australian English and Peruvian Spanish Listeners.声学特性可预测成年澳大利亚英语和秘鲁西班牙语听众对陌生荷兰元音的感知。
Front Psychol. 2017 Jan 27;8:52. doi: 10.3389/fpsyg.2017.00052. eCollection 2017.
7
Vowel and consonant confusions from spectrally manipulated stimuli designed to simulate poor cochlear implant electrode-neuron interfaces.旨在模拟人工耳蜗电极 - 神经元接口不佳情况的经频谱处理刺激所导致的元音和辅音混淆。
J Acoust Soc Am. 2016 Dec;140(6):4404. doi: 10.1121/1.4971420.
8
Influences of noise-interruption and information-bearing acoustic changes on understanding simulated electric-acoustic speech.噪声干扰和承载信息的声学变化对理解模拟电声语音的影响。
J Acoust Soc Am. 2016 Nov;140(5):3971. doi: 10.1121/1.4967445.
9
Nonlinear frequency compression: Influence of start frequency and input bandwidth on consonant and vowel recognition.非线性频率压缩:起始频率和输入带宽对辅音和元音识别的影响。
J Acoust Soc Am. 2016 Feb;139(2):938-57. doi: 10.1121/1.4941916.
10
The neural encoding of formant frequencies contributing to vowel identification in normal-hearing listeners.正常听力听众中有助于元音识别的共振峰频率的神经编码。
J Acoust Soc Am. 2016 Jan;139(1):1-11. doi: 10.1121/1.4931909.