• 제목/요약/키워드: Speech Tone

검색결과 201건 처리시간 0.019초

노인성 난청인의 음성특성에 관한 연구 (A study on speech analysis of person with presbycusis)

  • 이상민;송철규;우효창;이영묵;김원기
    • 대한의용생체공학회:학술대회논문집
    • /
    • 대한의용생체공학회 1997년도 추계학술대회
    • /
    • pp.67-70
    • /
    • 1997
  • In this paper, we evaluated the character of speech of hearing impaired person (HIP) who acquire his hearing loss after the youth. It is usually observed that severe HIP decreased not only speech perception but also vocalization. so there is a need for sensitive and quantitative measures or the assesment of the speech of the HIP to serve both diagnostic and prognosic purposes, 7 HIP and 12 normal hearing person(NHP) were studied with pure tone test and speaking test using word/sentence table which consists of vowel(a:), mono and two syllables and a sentence. we analyzed formant frequency, pitch, sound intensity, speech duration of HIP and NHP speech. According to the results, in the HIP's speech we find that formant frequency was shifted, first-formant prominence was reduced, the dynamic range of sound intensity was decreased, speech duration was prolonged. In the next, we expect the correlation between hearing and speech character of HIP is cleared through analysis of more acoustic parameters and precise selection of HIP group.

  • PDF

음성합성을 위한 C-ToBI기반의 중국어 운율 경계와 F0 contour 생성 (Chinese Prosody Generation Based on C-ToBI Representation for Text-to-Speech)

  • 김승원;정옥;이근배;김병창
    • 대한음성학회지:말소리
    • /
    • 제53호
    • /
    • pp.75-92
    • /
    • 2005
  • Prosody Generation Based on C-ToBI Representation for Text-to-SpeechSeungwon Kim, Yu Zheng, Gary Geunbae Lee, Byeongchang KimProsody modeling is critical in developing text-to-speech (TTS) systems where speech synthesis is used to automatically generate natural speech. In this paper, we present a prosody generation architecture based on Chinese Tone and Break Index (C-ToBI) representation. ToBI is a multi-tier representation system based on linguistic knowledge to transcribe events in an utterance. The TTS system which adopts ToBI as an intermediate representation is known to exhibit higher flexibility, modularity and domain/task portability compared with the direct prosody generation TTS systems. However, the cost of corpus preparation is very expensive for practical-level performance because the ToBI labeled corpus has been manually constructed by many prosody experts and normally requires a large amount of data for accurate statistical prosody modeling. This paper proposes a new method which transcribes the C-ToBI labels automatically in Chinese speech. We model Chinese prosody generation as a classification problem and apply conditional Maximum Entropy (ME) classification to this problem. We empirically verify the usefulness of various natural language and phonology features to make well-integrated features for ME framework.

  • PDF

The acquisition of boundary tones in spontaneous speech by Korean learners of English

  • Choe, Wook Kyung
    • 말소리와 음성과학
    • /
    • 제12권4호
    • /
    • pp.47-55
    • /
    • 2020
  • The current study was designed to investigate which type of phrase boundary tones high-intermediate Korean learners of English used in their spontaneous speech. These boundary tones were compared to those used in native speakers' spontaneous speech to examine whether the learners successfully acquired the use of boundary tones. To achieve this purpose, 10 Korean learners of English and four native speakers of English participated in the current study. The participants were asked to summarize the stories of short videos, and the tonal and the phrasing patterns of the obtained spontaneous speech were analyzed using Tone and Break Indices (ToBI) transcription conventions. The results indicated that both the native speakers and the Korean learners frequently marked their intonational phrase boundaries with high boundary tones. However, regarding the prosodic phrase positions within a sentence, Korean learners frequently used steep rising tones (i.e., H-H%) while native speakers used gradual rising tones (i.e., L-H%) for sentence-final intonational phrases. Overall, the findings suggested that high-intermediate Korean learners understood the forward-looking function of the high boundary tones and that they were able to make use of these tones to mark intonational phrases in their spontaneous speech.

표준 중국어의 경계억양에 관한 연구 (Study of Boundary Tone in Mandarin Chinese)

  • 손남호
    • 대한음성학회:학술대회논문집
    • /
    • 대한음성학회 2003년도 5월 학술대회지
    • /
    • pp.43-47
    • /
    • 2003
  • This paper is phonetic study of $F_{0}$ range and boundary tone in Mandarin Chinese. The production data from 6 Chinese speakers show that there are declination, pitch resetting and tonal variation of boundary tone. In declarative sentence, $F_{0}$ declines gradually over the utterance but mid-sentence boundary prevents $F_{0}$ of following syllable from declining because of pitch resetting. $F_{0}$ range of syllable is expanded before the mid- and final sentence boundaries. In interrogative one, $F_{0}$ ascends gradually over the utterance and mid-sentence boundary makes $F_{0}$ of following syllable rise more. $F_{0}$ range of sentence final syllable is expanded and $F_{0}$ contour shows rising curve.

  • PDF

한국어 운율구조와 관련한 모음 및 음절 길이 (On vowel and syllable duration related to prosodic structure in Korean)

  • 이숙향
    • 대한음성학회지:말소리
    • /
    • 제35_36호
    • /
    • pp.13-24
    • /
    • 1998
  • This study aims at examining the relationship between tonal events and their related vowel and syllable duration in Korean. Two things were investigated: one is to see if there is a hierarchical relationship in prosodic unit-final-lengthening and the other is to see if accentual phrase initial high tone syllable gets lengthened. Generally, higher prosodic units show larger degree of lengthening of the final vowel and also final syllable duration than the lower ones except for accentual phrase: Mean duration of utterance-final or intonational-phrase-final syllable(and its vowels) was longer than that of accentual-phrase-final or word-final syllable(and its vowels). However, mean duration of accentual phrase final syllable was shorter than that of word final syllable. Mean vowel duration of accentual phrase initial high tone syllable was shorter than that of any other prosodic unit. Its mean syllable duration, however, was longer than that of accentual-phrase-final or word-final syllable, indicating that strong consonants(fortis and aspirated) frequently appear in the accentual phrase initial position and this position is a prosodically strong position showing longer duration as well as high tone.

  • PDF

음성인식을 이용한 자동 호 분류 철도 예약 시스템 (A Train Ticket Reservation Aid System Using Automated Call Routing Technology Based on Speech Recognition)

  • 심유진;김재인;구명완
    • 대한음성학회지:말소리
    • /
    • 제52호
    • /
    • pp.161-169
    • /
    • 2004
  • This paper describes the automated call routing for train ticket reservation aid system based on speech recognition. We focus on the task of automatically routing telephone calls based on user's fluently spoken response instead of touch tone menus in an interactive voice response system. Vector-based call routing algorithm is investigated and mapping table for key term is suggested. Korail database collected by KT is used for call routing experiment. We evaluate call-classification experiments for transcribed text from Korail database. In case of small training data, an average call routing error reduction rate of 14% is observed when mapping table is used.

  • PDF

학령전기아동 관련 성인의 운율 특성 (The Prosodic Characteristics of Pre-school Age Children-Related Adults)

  • 김지원;성철재
    • 말소리와 음성과학
    • /
    • 제6권3호
    • /
    • pp.23-32
    • /
    • 2014
  • This study presents the prosodic characteristics of 'Motherese' and 'Teacherese (child care teacher and kindergarten teacher)'. 21 mothers and 24 teachers spoke to children in the child care center or kindergarten. Children are in their 4;00-6;11. Speech and articulation rate, number of accentual phrases (APs), number of intonational phrases (IPs), pitch-related factors (f0, pitch range, f0 standard deviation), and intonation slope (mean Absolute, f0, q-tone slope) were measured. 2 groups spoke 2 sentential types (interrogative_ alternative question, declarative_ coordinated sentence) in 2 situations (one accompanied with the children, the other done without children, but pretending as if they were in front of the children). The results indicate that teachers show more noticeable prosodic characteristics than mothers do.

비대칭 4 질량 성대 모델에 의한 쉰목소리 분석 (Hoarse Speech Analysis Using Dissymmetric Four-Mass Model of Vocal Cords)

  • 장강의;진혜방;최태영
    • 한국음향학회지
    • /
    • 제14권5호
    • /
    • pp.94-101
    • /
    • 1995
  • 본 논문에서는 쉰 목소리 메커니즘 분석을 위한 4질량 성대 모델을 제안하였다. 쉰 목소리가 성대의 병리학적 변화에 기인한다는 것과 성문 파형이 성대의 움직임 상태를 반영한다는 사실에서, 병든 성대를 비대칭 구조이고 4질량형으로 가정하였다. 정상 목소리와 쉰 목소리에 대한 모델 변수들과 성문 파형을 분석하여 모델 변수와 병리학 사이의 관계를 검토하였다. 실험 결과 쉰 목소리의 음향 특징과 병리학간의 관계를 밝힐 수 있었고 후두 질병 진단과 쉰 목소리의 음질 향상에도 본 논문에서 제안한 방법이 사용될 수 있음을 알았다.

  • PDF

잡음 환경에서 압신을 이용한 인공 와우 환자의 언어 인지 향상 시뮬레이션 연구 (A simulation study of speech perception enhancement for cochlear implant patients using companding in noisy environment)

  • 이영우;지윤상;이종실;김인영;김선일;홍성화;이상민
    • 대한전자공학회논문지SP
    • /
    • 제43권5호
    • /
    • pp.79-87
    • /
    • 2006
  • 본 연구에서 인공 와우 환자의 잡음 상황에서 음성 신호 강조와 잡음 제거를 위한 전 처리로서 companding strategy를 적용하고 이를 평가하였다. Companding은 인간의 청각 특성인 two tone suppression에 기반하며 이는 음성 스펙트럼 피크를 강화하고 배경 잡음을 감소시킨다. 하지만 companding은 잡음 제거와 스펙트럼 피크의 강화에 효과적인 반면, 제한된 채널의 수와 비선형 블록으로 인한 음성 정보 손실의 교환 특성을 가진다. 따라서 본 연구에서는 잡음 제거와 음성 정보 손실의 정도가 상대적인 두 companding 구조를 설계하여 개인마다 잡음 상황에서 언어 인지 특성차이에 따른 적절한 필터 뱅크를 도출하였으며, 낮은 신호 대 잡음 비 환경에서 인공 와우 환자의 언어 인지 향상을 위한 방법을 제시하였다. 제안된 알고리즘은 잡음 밴드 시뮬레이션을 이용하여 정상인 5명에게 평가되었다. 모든 피실험자에게서 효과적인 언어 인지의 향상이 관측되었고, 각 피실험자가 선호하는 필터 뱅크는 다르게 나타났다.

감정 음성의 국어 발화 말 경계성조 연구 (Research of Korean utterance-final boundary tones in Emotion speeches)

  • 박미영
    • 대한음성학회:학술대회논문집
    • /
    • 대한음성학회 2007년도 한국음성과학회 공동학술대회 발표논문집
    • /
    • pp.193-196
    • /
    • 2007
  • The purpose of this paper is to find boundary tone's characteristics in Korean emotion speeches. I mainly focus on investigating patterns and f0 values of boundary tones and f0 values in utterance final phrases.

  • PDF