Search | Korea Science

Hangeul detection method based on histogram and character structure in natural image (다양한 배경에서 히스토그램과 한글의 구조적 특징을 이용한 문자 검출 방법)

Pyo, Sung-Kook;Park, Young-Soo;Lee, Gang Seung;Lee, Sang-Hun
- Journal of the Korea Convergence Society
- /
- v.10 no.3
- /
- pp.15-22
- /
- 2019
In this paper, we proposed a Hangeul detection method using structural features of histogram, consonant, and vowel to solve the problem of Hangul which is separated and detected consonant and vowel The proposed method removes background by using DoG (Difference of Gaussian) to remove unnecessary noise in Hangul detection process. In the image with the background removed, we converted it to a binarized image using a cumulative histogram. Then, the horizontal position histogram was used to find the position of the character string, and character combination was performed using the vertical histogram in the found character image. However, words with a consonant vowel such as '가', '라' and '귀' are combined using a structural characteristic of characters because they are difficult to combine into one character. In this experiment, an image composed of alphabets with various backgrounds, an image composed of Korean characters, and an image mixed with alphabets and Hangul were tested. The detection rate of the proposed method is about 2% lower than that of the K-means and MSER character detection method, but it is about 5% higher than that of the character detection method including Hangul.
https://doi.org/10.15207/JKCS.2019.10.3.015 인용 PDF KSCI HTML

Multi-Strata Lexikon vs. Constraintranking: Degemination im Deutschen (다층어휘부와 어휘부 대 제약우위도)

Yu Si-Taek
- Koreanishche Zeitschrift fur Deutsche Sprachwissenschaft
- /
- v.1
- /
- pp.313-348
- /
- 1999
이 논문은 독일어의 겹자음회피현상을 설명함에 있어 어휘음운론에서의 분석이 보이는 문제들을 지적하고, 이 문제들이 제약에 바탕을 둔 이론에서는 어떻게 해결될 수 있는가를 보인다. 제약들간의 상호작용에서 특히 중요한 역할을 하는 것이 단일형태실현제약 (Uniform Exponence)으로서, 이 제약을 통해 독일어 동사의 현재시제, 단수, 2인칭 형태와 3인칭형태에서 나타나는 겹자음회피현상이 동사의 어형변화표 (Verbparadigma)와 밀접한 관련이 있음을 알 수 있다. 이는 규칙들을 통해 2인칭과 3인칭의 올바른 형태를 각각 개별적으로 찾아내는 어휘음운론의 분석과는 근본적으로 다르다. 왜냐하면 어휘음운론의 분석에 따를 때, 예를 들어 3인칭 동사 arbeitet에서 Schwa 모음의 삽입은 겹자음회피를 위해 일어난다고 설명되지만 겹자음이 없음에도 불구하고 Schwa 모음이 마찬가지로 삽입되는 2인칭동사 arbeitest는 설명되지 않기 때문이다. 이런 분석에서는 2인칭 형태와 3인칭 형태가 서로 아무런 관련 없이 각기 따로 존재하게된다. 이에 반해 단일형태제약은 이 두개의 형태를 동시에 비교하므로, 동사 굴절형태에서 마치 불필요한 것으로 보이는 모음삽입이나 자음탈락의 원인에 대해 이론적인 근거를 제시할 수 있다. 즉 2인칭 형태와 3인칭 형태는 보다 상위의 제약들이 막지 않는 한 서로 최대한 비슷한 형태를 가지려고 한다. 이 논문은 겹자음 회피를 위한 수단으로서 모음삽입이나 자음탈락은 오로지 이를 통해 동사의 어형변화표가 좋아질 때만 가능하다는 것을 보여줌으로써 규칙이론이 포착하지 못하고 있는 중요한 일반화를 제시하고 있다. 단일형태 실현제약의 중요성은 접두사 in- 과 un- 이 어간과 결합할 때 보이는 대조를 통해서도 확인된다. 여기서도 어휘음운론의 다층어휘부 구조에 의한 설명이 갖는 문제점이 제약들간의 상호작용을 통해 해결될 수 있음을 알 수 있다.VII-1 및 VII-2공의 3600 m 하부층은 건성 가스 생성 단계에까지 도달한 것으로 나타났다. JDZ VII-1, VII-2 시추공의 3500 m 하위 구간의 올리고세 퇴적층에서 유기물 함량 및 수소 지수가 급격히 감소하는 것은 매몰 심도가 깊어지면서 유기물이 열 분해되어 이미 탄화수소를 생성한 것으로 해석된다. JDZ VII-1 및 VII-2 시추공의 가스징후 및 길소나이트 (gilsonite)는 탄화수소가 생성되어 이동한 흔적을 시사한다.을 해석할 수 있음을 보여주는 것으로 평가된다. 다만 PLAYMAKER2가 보다 신뢰할 만한 퇴적환경 해석을 위한 전문가 시스템으로 구축되기 위해서는 향후 많은 퇴적학 전문가들이 추가로 참여하여 기존 규칙들을 재검증하고 새로운 규칙들을 첨가함으로써 보디 세련된 지식베이스를 개발하여야 할 것으로 판단된다.이며 세 개의 산소가 이루는 평면에서 $1.68{\AA}$ 소다라이트내로 이동하여 위치한다. 32개의 $Tl^{+}$ 이온은 결정학적 자리 II에 존재하고 있으며 산소와의 결합거리를 $2.70(1){\AA}$을 유지하면서 큰 동공속으로 $1.48{\AA}$ 이동하여 위치한다. 약 18개의 $Tl^+$ 이온은 결정학적 자리III에, 또 다른 10개의 $Tl^+$ 이온은 결정학적 자리III'에 존재하고 골조 산소와 각각 $2.86(2){\AA},\;2.96(4){\AA}$의 결합거리를 이룬다.
PDF

The syllable recovery rule-base system for the post-processing of a continuous speech recognition (연속음성인식 후처리를 위한 음절 복원 rule-base시스템)

Park, Mi-Seong;Kim, Mi-Jin;Lee, Mun-Hui;Choi, Jae-Hyeok;Lee, Sang-Jo
- Annual Conference on Human and Language Technology
- /
- 1998.10c
- /
- pp.379-385
- /
- 1998
한국어가 연속적으로 발음될 때 여러 가지 음운 변동현상이 일어난다. 이것은 한국어 연속음성 인식을 어렵게 하는 주요 요인 중의 한가지이다. 본 논문은 음운변동현상이 반영된 음성 인식 문자열을 규칙에 의거하여 text 기반 문자열로 다시 복원시키고 복원 결과 후보들을 형태소 분석하여 유용한 문자열만을 최종 결과로 생성하게 하는 시스템을 구성하였다. 복원은 4가지 rule 즉, 음절 경계 종성 초성 복원 rule, 모음처리 복원 rule, 끝음절 중성 복원 rule, 한 음절처리 rule에 따라 이루어진다. 규칙 적용 과정중에 효과적인 복원을 위해 x-clustering정보를 정의 하여 사용하고, 형태소 분석기에 입력될 복원 후보수를 제한하기 위해 postfix음절 빈도정보를 구하여 사용한다.
PDF

The syllable recovrey rule-based system and the application of a morphological analysis method for the post-processing of a continuous speech recognition (연속음성인식 후처리를 위한 음절 복원 rule-based 시스템과 형태소분석기법의 적용)

박미성;김미진;김계성;최재혁;이상조
- Journal of the Korean Institute of Telematics and Electronics C
- /
- v.36C no.3
- /
- pp.47-56
- /
- 1999
Various phonological alteration occurs when we pronounce continuously in korean. This phonological alteration is one of the major reasons which make the speech recognition of korean difficult. This paper presents a rule-based system which converts a speech recognition character string to a text-based character string. The recovery results are morphologically analyzed and only a correct text string is generated. Recovery is executed according to four kinds of rules, i.e., a syllable boundary final-consonant initial-consonant recovery rule, a vowel-process recovery rule, a last syllable final-consonant recovery rule and a monosyllable process rule. We use a x-clustering information for an efficient recovery and use a postfix-syllable frequency information for restricting recovery candidates to enter morphological analyzer. Because this system is a rule-based system, it doesn't necessitate a large pronouncing dictionary or a phoneme dictionary and the advantage of this system is that we can use the being text based morphological analyzer.
PDF

SHRT : New Method of URL Shortening including Relative Word of Target URL (SHRT : 유사 단어를 활용한 URL 단축 기법)

Yoon, Soojin;Park, Jeongeun;Choi, Changkuk;Kim, Seungjoo
- The Journal of Korean Institute of Communications and Information Sciences
- /
- v.38B no.6
- /
- pp.473-484
- /
- 2013
Shorten URL service is the method of using short URL instead of long URL, it redirect short url to long URL. While the users of microblog increased rapidly, as the creating and usage of shorten URL is convenient, shorten url became common under the limited length of writing on microblog. E-mail, SMS and books use shorten URL well, because of its simplicity. But, there is no relativeness between the most of shorten URLs and their target URLs, user can not expect the target URL. To cover this problem, there is attempts such as changing the shorten URL service name, inserting the information of website into shorten URL, and the usage of shortcode of physical address. However, each ones has the limits, so these are the trouble of automation, relatively long address, and the narrowness of applicable targets. SHRT is complementary to the attempts, as getting the idea from the writing system of Arabic. Though the writing system of Arabic has no vowel alphabet, Arabs have no difficult to understand their writing. This paper proposes SHRT, new method of URL Shortening. SHRT makes user guess the target URL using Relative word of the lowest domain of target URL without vowels.
https://doi.org/10.7840/kics.2013.38B.6.473 인용 PDF KSCI

Speech analysis using the Robust Time-Weighted Kalman filtering (시간가중치의 로버스트 칼만필터를 이용한 음성분석)

최홍섭;안수길
- The Journal of the Acoustical Society of Korea
- /
- v.11 no.1E
- /
- pp.73-78
- /
- 1992
시벼형 신호인 음성 신호의 분석에 칼만필터를 이용하였다. 일반적인 음성 분석은 프레임단위의 처리방법인 선형 예측 부호화 기법을 주로 이용하지만 음성의 시변 특성을 파악하는데에는 적절하지 못 하다. 따라서 순차적인 추정기법으로 많이 이용되는 칼만 필터를 음성 분석에 적용하였다. 또한 음성과 같은 시변신호에서는 과거 신호의 잡음의 분산값에 적당한 가중치를 부가하므로써 과거의 신호에 의해 서 현재의 추정값에 미치는 영향을 줄였으며 이를 음성의 천이 구간에서의 파라메타 추정에 사용하였 다. 그리고 음성신호 모델에서 생기는 모델링 오차는 일반적으로 백색 가우시안 잡음으로 가정하고 있 으나 이는 자음과 같은 무성음에서 특징 파라메타 푸정에는 오차가 적지만 모음등의 유성음에서는 음성 발생시의 여기신호인 펄스열에 의해서 많은 모델링 오차를 생기게 한다. 따라서 모델링 오차신호는 Non-Gaussian 확률분포로 가정한 후 로버스트 칼만 필터를 사용하여 합성으멩 대해 특징 파라메터를 추출하였다.
PDF

An Experimental Studies on Vowel Duration Differences before Consonant Clusters and unreleased stops of coda-position (영어 어말 자음군 구성에 따른 선행모음 길이 변화 및 어말 자음 비파열 현상에 대한 실험음성학적 연구 -무성 폐쇄음을 중심으로-)

Shin Dong-Jin
- Proceedings of the KSPS conference
- /
- 2006.05a
- /
- pp.55-58
- /
- 2006
The aim of this paper is to investigate the effects of postvocalic consonant cluster (Contrasting nasal-stops consonant with stops) on vowel duration. In particular we focused on the rate of vowel duration in their words. (Experimental I ) and the tendency of unreleased voiceless stops at the end of the words.(Experimental II). The result of experimental I showed that the rate of vowel duration which is preceding single voiceless stops are significantly longer than those preceding nasal-stops counterparts and the percentage of English native speakers was longer than those of Korean leaners of English Experiment II indicated that the tendency of unreleased stop consonants occurred more frequently on single voiceless stops than nasal-stop clusters and Korean learners of English were more frequently produced the unreleased stops than English natives.
PDF

A Study on 7-Connected Digits Speech Recognition using SCHMM (SCHMM 기반 7연속 숫자음 인식에 관한 연구)

Kim Se Yong;Jung Hui Seok;Kang Chul Ho
- Proceedings of the Acoustical Society of Korea Conference
- /
- spring
- /
- pp.127-130
- /
- 2002
본 연구에서는 우리말 연속 숫자음 인식에서 본래의 숫자음을 변이 시키는 주된 요인인 연음현상에 대한 인식을 높이기 위해 별도의 연음부분의 레퍼런스를 작성하여 매칭 시키는 방식을 제안한다 또한 단모음으로 이루어진 /2/와 /5/의 연속된 음에 대하여도 레퍼런스를 작성하였다. 제안한 방식에 의하여 전체적으로 $1.4\%$정도 인식률이 상승됨을 볼 수 있다. 특히 발성 목록중 /82/, /62/, /31/, /15/, /75/ 등의 연음과 /226/, /755/등과 같이 모음의 연속된 발성이 포함된 숫자 열에서 제안된 방식이 인식률에 영향을 미치는 것을 볼 수가 있었다. 이는 연음에서 발생하는 오류가 연속 숫자음에 많은 영향을 미치는 것을 알 수 있다. 그 외에 /22/, /55/등과 같이 단모음으로 이루어진 숫자음의 연속 발성 또한 인식률을 저하시키는데 한 요인으로 작용함으로서 이에 대한 레퍼런스도 작성하여 인식률이 상승되는 것을 볼 수 있었다.
PDF

계층적 신경망을 이용한 자소인식에 기초한 Off-Line 필기체 한글인식 : 자소간 섭동체거를 위한 High-Level Constraint 회로의 설계

장주석;김명원;임채덕;송윤선
- Information and Communications Magazine
- /
- v.9 no.11
- /
- pp.34-36
- /
- 1992
여러 개의 문자(혹은 여러 개의 자소로 구성된 한개의 문자)를 인식할때에는 문자(혹은 자소) 상호간에 영향을 미쳐서 오인식이 발생할 가능성이 높다. 개개의 숫자인식에 기초한 숫자열 인식이나, 개개의 자소인식을 바탕으로한 필기체 한글인식이 그 좋은 보기일 것이다. 예를 들어 단순한 한글 '그'를 Neocognitron으로 인식한다고 생각해 보자, 조합 가능한 글자를 모두 기억시키려면 방대한 규모의 회로가 필요하므로 현실적으로 불가능하다. 따라서 기본 자소(자음 14개, 모음 10개)를 인식하도록 학습시키고 이를 바탕으로 한글을 인식하는 것이 효율적이다. 이때, 회로의 각 세포가 보는 receptive field가 유한하여 '?'의 끝 세로부분 'I'가 '?'에 영향을 미쳐서 '?'로 인식된다 즉, 자소간의 섭동에 의해 '그'가 '고'로 인식되는 것이다. 이와같은 예는 '니'가 '넉'으로, '41'이 '4H'로 인식되는 등 매우 많지만 그 해결에 대한 연구는 거의 없다. 이 논문에서는 필기체 한글 자소를 인식하는 Necognitron외에 자소간의 섭동현상을 제거하기 위한 high-level constraint 회로를 Lotka-Volterra동역학에 기초하여 설계하였다. 이로써 off-line필기체 한글인식을 보다 효과적으로 할 수 있음을 컴퓨터 시뮬레이션으로 보인다.
PDF

Development of a Lipsync Algorithm Based on Audio-visual Corpus (시청각 코퍼스 기반의 립싱크 알고리듬 개발)

김진영;하영민;이화숙
- The Journal of the Acoustical Society of Korea
- /
- v.20 no.3
- /
- pp.63-69
- /
- 2001
A corpus-based lip sync algorithm for synthesizing natural face animation is proposed in this paper. To get the lip parameters, some marks were attached some marks to the speaker's face, and the marks' positions were extracted with some Image processing methods. Also, the spoken utterances were labeled with HTK and prosodic information (duration, pitch and intensity) were analyzed. An audio-visual corpus was constructed by combining the speech and image information. The basic unit used in our approach is syllable unit. Based on this Audio-visual corpus, lip information represented by mark's positions was synthesized. That is. the best syllable units are selected from the audio-visual corpus and each visual information of selected syllable units are concatenated. There are two processes to obtain the best units. One is to select the N-best candidates for each syllable. The other is to select the best smooth unit sequences, which is done by Viterbi decoding algorithm. For these process, the two distance proposed between syllable units. They are a phonetic environment distance measure and a prosody distance measure. Computer simulation results showed that our proposed algorithm had good performances. Especially, it was shown that pitch and intensity information is also important as like duration information in lip sync.
PDF

Search Result 25, Processing Time 0.025 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)