Automatic recognition of gemination in Japanese motivated by perceptual experiments
スポンサーリンク
概要
- 論文の詳細を見る
For Japanese speech processing, being able to automatically recognize between geminate and singleton consonants can have many benefits. In standard recognition methods, hidden Markov Models (HMMs) are used. However, HMMs are not good at differentiating between items that are distinguished primarily by temporal differences rather than spectral differences. Also, gemination depends on the length of the sounds surrounding the consonant. Because of this, we propose the construction of a method that automatically distinguishes geminates from singletons and takes these factors into account. In order to do this, it is necessary to determine which surrounding sounds are cues and what the mechanism of human recognition is. For this, we conduct perceptual experiments to examine the relationship between surrounding sounds and primary cues. Then, using these results, we design a method that can automatically recognize gemination. We test this method on two datasets including a speaking rate database. The results attained well-outperform the HMM-based method and overall outperform the case when only the primary cue is used for recognition as well as show more robustness against speaking rate.
- 一般社団法人 日本音響学会の論文
一般社団法人 日本音響学会 | 論文
- How large is the individual difference in hearing sensitivity?: Establishment of ISO 28961 on the statistical distribution of hearing thresholds of otologically normal young persons
- Applying generation process model constraint to fundamental frequency contours generated by hidden-Markov-model-based speech synthesis
- Vocal cord vibration in the production of consonants. Observation by means of high-speed digital imaging using a fiberscope.:Observation by means of high-speed digital imaging using a fiberscope
- The early reflections of the impulse response in an auditorium.
- Multiple reflections between rigid plane panels.