A study on the change of prosodic units by speech rate and frequency of turn-taking (발화 속도와 말차례 교체 빈도에 따른 운율 단위 변화에 관한 연구)

  • Won, Yugwon
    • Phonetics and Speech Sciences
    • v.14 no.2
    • pp.29-38
    • 2022
  • This study aimed to analyze the speech appearing in the National Institute of Korean Language's Daily Conversation Speech Corpus (2020) and reveal how the speech rate and the frequency of turn-taking affect the change in prosody units. The analysis results showed a positive correlation between intonation phrase, word phrase frequency, and speaking duration as the speech speed increased; however, the correlation was low, and the suitability of the regression model of the speech rate was 3%-11%, which was weak in explanatory power. There was a significant difference in the mean speech rate according to the frequency of the turn-taking, and the speech rate decreased as the frequency of the turn-taking increased. In addition, as the frequency of turn-taking increased, the frequency of intonation phrases, the frequency of word phrases, and the speaking duration decreased; there was a high negative correlation. The suitability of the regression model of the turn-taking frequency was calculated as 27%-32%. The frequency of turn-taking functions as a factor in changing the speech rate and prosodic units. It is presumed that this can be influenced by the disfluency of the dialogue, the characteristics of turn-taking, and the active interaction between the speakers.

The Rule of Korean Pitch Variation for a Natural Synthetic Female Voice (자연스러운 여성 합성음을 위한 한국어의 피치 변화 법칙)

  • Kim, Chung-Won;Park, Dae-Duck;Kim, Boh-Hyun;Kwon, Cheol-Hong
    • The Journal of the Acoustical Society of Korea
    • v.15 no.6
    • pp.26-32
    • 1996
  • In this paper we make a rule of pitch variation for a natural synthetic female voice. Intonation phrase, which is the basic unit the rule is applied to, mostly consists of a syllable or syllables. The pitch values of the first, second, and final syllables make up the pitch contour of the intonation phrase. Those of the first and second syllable are determined by the initial consonants of the respective syllables, and that of the final syllable by the type of the function word. There are two kinds of boundaries between intonation phrases. One is a boundary with pause, and the other is a boundary without pause. The pitch contour of the intonation phrase with the boundary phenomena determines the pitch pattern of a sentence.

Acoustic Phonetic Study about Focus Realization of wh-word Questions in Korean (국어 의문사${\cdot}$부정사 의문문의 초점 실현에 대한 음향음성학적 연구)

  • Park Mi-young;Ahn Byoung-seob
    • Proceedings of the Acoustical Society of Korea Conference
    • spring
    • pp.289-292
    • 2002
  • 국어에서 wh-단어가 포함된 의문사 의문문과 부정사 의문문은 통사적으로 같은 구조를 가지지만 의미적으로는 중의 관계에 있다. 그러나 두 의문문은 문장으로 발화될 때 음성적으로 서로 다른 여러 가지 운율 특징의 차이를 보여줌으로써, 발화 차원에서는 더 이상 중의 관계를 유지하지 않는다. 본고에서는 이러한 중의성의 해소는 두 의문문의 초점이 달리 실현되기 때문이라고 본다. 기존의 연구에서는 두 가지 의문문의 억양 연구를 초점의 작용 범위와 문말 억양의 차이, 강세구 형성의 유형을 중심으로 고찰하였다 .그리고 의문사와 부정사의 의미는, 이에 후행하는 서술어와 형성하는 강세구 유형에서 우선적으로 그 의미가 구분될 수 있다고 보았다. 그러나, 본고에서는 국어의 wh-단어가 초점으로서 작용하는 운율적 돋들림을 좀더 다양한 환경에서 실험하였다. 그리고 의문사${\cdot}$부정사와 후행하는 언어단위의 강세구 형성(accentual phrasing) 유형, 의문사${\cdot}$부정사 의문문 전체 문장 억양의 실현 양상, wh-단어 자체의 음의 높낮이(pitch contour) 실현 유형, 문말 억양(boundary tone)에서 음의 높낮이를 대상으로 분석하였다.

A Pre-Selection of Candidate Units Using Accentual Characteristic In a Unit Selection Based Japanese TTS System (일본어 악센트 특징을 이용한 합성단위 선택 기반 일본어 TTS의 후보 합성단위의 사전선택 방법)

  • Na, Deok-Su;Min, So-Yeon;Lee, Kwang-Hyoung;Lee, Jong-Seok;Bae, Myung-Jin
    • The Journal of the Acoustical Society of Korea
    • v.26 no.4
    • pp.159-165
    • 2007
  • In this paper, we propose a new pre-selection of candidate units that is suitable for the unit selection based Japanese TTS system. General pre-selection method performed by calculating a context-dependent cost within IP (Intonation Phrase). Different from other languages, however. Japanese has an accent represented as the height of a relative pitch, and several words form a single accentual phrase. Also. the prosody in Japanese changes in accentual phrase units. By reflecting such prosodic change in pre-selection. the qualify of synthesized speech can be improved. Furthermore, by calculating a context-dependent cost within accentual phrase, synthesis speed can be improved than calculating within intonation phrase. The proposed method defines AP. analyzes AP in context and performs pre-selection using accentual phrase matching which calculates CCL (connected context length) of the Phoneme's candidates that should be synthesized in each accentual phrase. The baseline system used in the proposed method is VoiceText, which is a synthesizer of Voiceware. Evaluations were made on perceptual error (intonation error, concatenation mismatch error) and synthesis time. Experimental result showed that the proposed method improved the qualify of synthesized speech. as well as shortened the synthesis time.

On a Template Extraction of phrase unit by Pitch Searching (피치 검색에 의한 Phrase 단위의 Template 추출에 관한 연구)

  • Kim JongKuk;Bae MyungJin
    • Proceedings of the Acoustical Society of Korea Conference
    • autumn
    • pp.77-80
    • 2004
  • 원화자로부터 목표 화자의 음성으로 변환을 위해서는 음운 및 피치변환이 이루어져야 한다. 원 음성과 목표 음성 신호 사이에 따른 발성길이, 크기 및 피치 등의 운율 특성은 화자의 개인성 및 발성문장의 의도를 나타내는 주요 역할을 한다. 본 논문에서는 음성 변환을 수행하기 위하여 발성된 음성의 강세구(phrase)단위의 피치 검출을 통하여 템플릿을 추출하는 방법을 제안한다. 우선 한국어의 운율구에 대한 정보가 필요한 것인지, 한국어는 어떤 운율 구조를 갖는지에 대하여 알아본다. 마지막으로 어떻게 연속음성으로부터 한국어에 적당한 운율구 단위를 나눌 것인지, 즉 자동 세그멘테이션 및 레이블링에 대하여 분석한다. 또한 논문에서는 한국어 문장음성의 운율구를 강세구와 억양구로 나누고 육안으로 표시한 운율구 단위를 기준으로 이 운율구 단위에 적합한 특징을 추출하여 패턴을 작성한다.

On the relationship between the phonetic realizations of the allophones of the Korean liquid /l/ and their prosodic status (한국에 유음 /l/의 변이음들의 음성적 실현과 운율적 위상과의 상관관계에 관하여)

  • 이숙향
    • The Journal of the Acoustical Society of Korea
    • v.18 no.7
    • pp.85-91
    • 1999
  • The purpose of this study is to investigate phonetic realization of flap [r], one of the allophones of Korean /l/. Phonetic realization of a segment is affected by not only its neighboring segments but also its prosodic position in an utterance. This study examined how various prosodic positions affect the phonetic realization of [r]. Effects of the four prosodic positions on the phonetic realization of [r] were examined: utterance initial, Intonation Phrase initial, Accentual Phrase initial, and Accentual Medial positions. Word positional effect was also examined: word initial, medial, and final positions. Acoustic and statistical analyses showed that flap [r] was realized in a variety of phonetic forms: from sonorant(the most reduced form) to short stop(the least reduced form). It was shown that generally. word-initial position is stronger than word-medial position. It was also shown that in many cases, utterance-initial position and intonation-phrase-initial position are stronger than accentual-phrase-initial and accentual-phrase-medial positions. Sonorants were observed more often in the prosodically weaker portions. VOT duration was also shorter in accentual-phrase-initial and accentual-phrase-medial positions.

Automatic Detection of Intonational and Accentual Phrases in Korean Standard Continuous Speech (한국 표준어 연속음성에서의 억양구와 강세구 자동 검출)

  • Lee, Ki-Young;Song, Min-Suck
    • Speech Sciences
    • v.7 no.2
    • pp.209-224
    • 2000
  • This paper proposes an automatic detection method of intonational and accentual phrases in Korean standard continuous speech. We use the pause over 150 msec for detecting intonational phrases, and extract accentual phrases from the intonational phrases by analyzing syllables and pitch contours. The speech data for the experiment are composed of seven male voices and two female voices which read the texts of the fable 'the ant and the grasshopper' and a newspaper article 'manmulsang' in normal speed and in Korean standard variation. The results of the experiment shows that the detection rate of intonational phrases is 95% on the average and that of accentual phrases is 73%. This detection rate implies that we can segment the continuous speech into smaller units(i.e. prosodic phrases) by using the prosodic information and so the objects of speech recognition can narrow down to words or phrases in continuous speech.

Correlation between tonal events and their acoustic duration (한국어 성조 이벤트와 음향적 길이)

  • 이숙향
    • Proceedings of the Acoustical Society of Korea Conference
    • 1998.06c
    • pp.383-386
    • 1998
  한국어의 운율구조는 발화문장(utterance), 억양구(intonational phrase), 악센트구(accentual phrase), 음운적 어절(phonological word), 음절(syllable) 순의 계층적 구조를 가지고 있다. 본 연구에서는 운율구조의 각 층에서 성조 이벤트가 얹혀지는 음절이나 또는 각 층의 운율단위말의 음절의 음향적 길이를 측정함으로써 첫째, 운율단위말의 음절의 음향적 길이 또한 계층적 순위를 보이는지 둘째, 성조 이벤트(tonal event)와 음향적 길이 사이에 높은 상관관계를 보이는지 보고자 한다. 즉, 두 가지 측면에서 길이비교가 수행되었는데 하나는 언어 보편적 현상으로 알려진 구말 장음화 현상으로써 각 층 운율적 단위의 마지막 음절의 모음 길이 비교이며 다른 하나는 억양구초 고성조가 실현되는 음절의 모음과 어절 내 모음, 그리고 고성조가 실현되는 억양구말 음절의 모음간의 길이 비교이다. 남녀 각각 200문장의 각 분절음과 운율분석을 한 후 길이에 대한 일원분산분석 실시 결과 억양구말은 악센트구말 보다 길었으나 악센트구말은 어절말과 차이를 보이지 않거나 남자 화자의 경우 오히려 짧게 나타났다. 그리고 남자화자의 경우 악센트구초 고성자가 얹혀지는 음절의 길이는 어절 내 어절말 음절을 제외한 그 외 음절과 화자에 따라 큰 차이를 보이지 않거나 그보다 조금 짧게 실현되는 것으로 나타났다. 위의 결과는 첫째, 단위말 음절 모음의 장음화는 운율적 구조의 층위에 일대일 대응을 보이지 않는 것으로 해석되며 둘째, 성조 이벤트와 그것이 실현되는 분절음의 음향적 길이와는 큰 상관관계를 보이지 않는 것으로 해석될 수 있겠다. 그러나 이러한 일반화에 대한 충분한 근거 제공을 위해서는 해당음절의 모음 길이 뿐만 아니라 초성자음의 길이간의 비교와 음절자체의 길이 비교 또한 필요한 것이며 모음길이에 대한 선행자음의 분절음적 영향 고려가 수반되어야 할 것으로 보인다.

A Comparative Study of Intonation Phrase Boundary Tones of Korean Produced by Korean Speakers and Chinese Speakers in the Reading of Korean Text (중국인 학습자들의 한국어 억양구 경계톤 실현 양상)

  • Yune, Young-Sook
    • Phonetics and Speech Sciences
    • v.2 no.4
    • pp.39-49
    • 2010
  • The purpose of this paper is to examine how Chinese speakers realize Korean intonation phrase (IP) boundary tones in the reading of a Korean text. Korean IP boundary tones play various roles in speech communication. They indicate prosodic constituents' boundaries while simultaneously performing pragmatic and grammatical functions. In order to express and understand Korean utterances correctly, it is necessary to understand the Korean IP boundary tone system. To investigate the IP boundary tone produced by Chinese speakers, we have specifically examined the type of boundary tones, the degree of internal pitch modulation of boundary tones, and the pitch difference between penultimate syllables and boundary tones. The results of each analysis were compared to the IP boundary tones produced by Korean native speakers. The results show that IP boundary tones were realized higher than penultimate syllables.

A Study on the Lexicalization of {Geuraegajigo} Based on the Spontaneous Speech Corpus (자유 발화 자료에 나타난 {그래가지고}의 접속 부사화)

  • Ha, Youngwoo;Shin, Jiyoung
    • Korean Linguistics
    • /
    • /
    • /
  • The aim of this paper is to study the morphemization of {Geuraegajigo} based on a spontaneous speech corpus. For this purpose, the distributions, the semantic functions, and the intonational phrase pattterns of the connective {Geuraegajigo} have been analyzed based on the corpus. The results are as follow; at first, coalescence that comes with a morphemization process was found, resulting in many variations. Secondly, there are three functions of it: [Direct/Indirect interrelationship], [Enumerate conjunction], and [Discourse marker]. And this semantic/functional diversity has many similarities with conjunctive adverbs. Lastly, intonational phrase patterns of {Geuraegajigo} accord with those of conjunctive adverbs. Especially, the discourse strategic IP pattern is connected with the short variation type. In conclusion, {Geuraegajigo} has finished turning into a conjunctive adverb through morphemization.