• Title/Summary/Keyword: Recognition unit

Search Result 517, Processing Time 0.022 seconds

A Study on Speech Recognition based on Phoneme for Korean Subway Station Names (한국의 지하철역명을 위한 음소 기반의 음성인식에 관한 연구)

  • Kim, Beom-Seung;Kim, Soon-Hyob
    • Journal of the Korean Society for Railway
    • /
    • v.14 no.3
    • /
    • pp.228-233
    • /
    • 2011
  • This paper presented the method about the Implementation of Speech Recognition based on phoneme considering the phonological characteristic for Korean Subway Station Names. The Pronunciation dictionary considering PLU set and phonological variations with four Case in order to select the optimum PLU used for Speech Recognition based on phoneme for Korean Subway Station Names was comprised and the recognition rate was estimated. In the case of the applied PLU, we could know the optimum recognition rate(97.74%) be shown in the triphone model in case of considering the recognition unit division of the initial consonant and final consonant and phonological variations.

A Study on Korean Allophone Recognition Using Hierarchical Time-Delay Neural Network (계층구조 시간지연 신경망을 이용한 한국어 변이음 인식에 관한 연구)

  • 김수일;임해창
    • Journal of the Korean Institute of Telematics and Electronics B
    • /
    • v.32B no.1
    • /
    • pp.171-179
    • /
    • 1995
  • In many continuous speech recognition systems, phoneme is used as a basic recognition unit However, the coarticulation generated among neighboring phonemes makes difficult to recognize phonemes consistently. This paper proposes allophone as an alternative recognition unit. We have classified each phoneme into three different allophone groups by the location of phoneme within a syllable. For a recognition algorithm, time-delay neural network(TDNN) has been designed. To recognize all Korean allophones, TDNNs are constructed in modular fashion according to acoustic-phonetic features (e.g. voiced/unvoiced, the location of phoneme within a word). Each TDNN is trained independently, and then they are integrated hierarchically into a whole speech recognition system. In this study, we have experimented Korean plosives with phoneme-based recognition system and allophone-based recognition system. Experimental results show that allophone-based recognition is much less affected by the coarticulation.

  • PDF

Machine Printed Character Recognition Based on the Combination of Recognition Units Using Multiple Neural Networks (다중 신경망을 이용한 인식단위 결합 기반의 인쇄체 문자인식)

  • Lim, Kil-Taek;Kim, Ho-Yon;Nam, Yun-Seok
    • The KIPS Transactions:PartB
    • /
    • v.10B no.7
    • /
    • pp.777-784
    • /
    • 2003
  • In this Paper. we propose a recognition method of machine printed characters based on the combination of recognition units using multiple neural networks. In our recognition method, the input character is classified into one of 7 character types among which the first 6 types are for Hangul character and the last type is for non-Hangul characters. Hangul characters are recognized by several MLP (multilayer perceptron) neural networks through two stages. In the first stage, we divide Hangul character image into two or three recognition units (HRU : Hangul recognition unit) according to the combination fashion of graphemes. Each recognition unit composed of one or two graphemes is recognized by an MLP neural network with an input feature vector of pixel direction angles. In the second stage, the recognition aspect features of the HRU MLP recognizers in the first stage are extracted and forwarded to a subsequent MLP by which final recognition result is obtained. For the recognition of non-Hangul characters, a single MLP is employed. The recognition experiments had been performed on the character image database collected from 50,000 real letter envelope images. The experimental results have demonstrated the superiority of the proposed method.

Education Equipment and Its Application for Indoor Position Recognition Using Inertial Measurement Unit Sensor (IMU센서를 이용한 실내 위치 인식 교육용 장비 및 응용)

  • Seo, Bo-In;Yu, YunSeop
    • Journal of Practical Engineering Education
    • /
    • v.10 no.2
    • /
    • pp.119-124
    • /
    • 2018
  • Educational equipment that enables the user or device to recognize the indoor position by using the acceleration and angular velocity of the IMU (Inertial Measurement Unit) sensor is introduced. With this educational equipment, various position recognition and tracking algorithms can be learned and creative engineering design works can be realized. The data value of the IMU sensor is transmitted to the MCU (microcontroller unit) through $I^2C$ (Inter-Integrated Circuit), and the indoor position recognition algorithm is applied by processing the data value through the filter and numerical method. It is then designed to use wireless communication to send and receive processed values and to be recognized by the user. As an example using this equipament, the case of "Implementation and recognition of virtual position using computation of moving direction and distance using IMU sensor" is introduced, and various creative engineering design application is discussed.

A Development of Underwater Sound Signal Recognition Algorithm for Acoustic Releaser in the Seafloor (심해저용 원격 착탈 시스템 제어를 위한 수중음향신호 인식 알고리즘의 개발)

  • 김영진;우종식;조영준;허경무
    • Journal of Institute of Control, Robotics and Systems
    • /
    • v.10 no.5
    • /
    • pp.421-427
    • /
    • 2004
  • In order to exploit underwater resources successfully, the first step would be a marine environmental research and exploration in the seafloor. Generally one sets up a long-term underwater experimental unit in the seafloor and retrieves the unit later after a certain period time. Essential to these applications is the reliable teleoperation and telemetering of the unit. In this paper we presents a robust underwater sound recognition algorithm by which we can identify the sound signal without the influence of disturbances due to underwater environmental changes. The proposed method provides a means suitable for the acoustic releaser which requires low power dissipation and long-time underwater operation. We demonstrate its ability of securing stability and fast sound recognition through simulation methods.

A Study on Korean Connected Digit Recognizer Based on Semi-syllable and Post-processing (반음절기반의 한국어 연속숫자음인식과 그 후처리에 대한 연구)

  • Jeong, Jae-Boo;Chung, Hoon;Chung, Ik-Joo
    • Speech Sciences
    • /
    • v.8 no.4
    • /
    • pp.1-15
    • /
    • 2001
  • This paper describes the effect of new recognition unit, a unit based on semisyllable, and its post processing method. A recognition unit based on semi-syllable expresses Korean connected digit's coarticulation effect. An existing method using semi-syllable limits next models, derived from current recognized models, to make complete connected digit sequence. However, this paper uses a new method to make complete connected digit sequence. The new post-processing method recognizes isolated digit words which include digits sequence from the digit combinations being able to occur from current recognized semi-syllable sequence. This method gives an improved accuracy rate than that of existing method. This new post processing provides two advantages. 1) It corrects current mis-recognized semi-syllable unit. 2) When people say each digit, they say it without regard to saying duration.

  • PDF

An Action Unit co-occurrence constraint 3DCNN based Action Unit recognition approach

  • Jia, Xibin;Li, Weiting;Wang, Yuechen;Hong, SungChan;Su, Xing
    • KSII Transactions on Internet and Information Systems (TIIS)
    • /
    • v.14 no.3
    • /
    • pp.924-942
    • /
    • 2020
  • The facial expression is diverse and various among persons due to the impact of the psychology factor. Whilst the facial action is comparatively steady because of the fixedness of the anatomic structure. Therefore, to improve performance of the action unit recognition will facilitate the facial expression recognition and provide profound basis for the mental state analysis, etc. However, it still a challenge job and recognition accuracy rate is limited, because the muscle movements around the face are tiny and the facial actions are not obvious accordingly. Taking account of the moving of muscles impact each other when person express their emotion, we propose to make full use of co-occurrence relationship among action units (AUs) in this paper. Considering the dynamic characteristic of AUs as well, we adopt the 3D Convolutional Neural Network(3DCNN) as base framework and proposed to recognize multiple action units around brows, nose and mouth specially contributing in the emotion expression with putting their co-occurrence relationships as constrain. The experiments have been conducted on a typical public dataset CASME and its variant CASME2 dataset. The experiment results show that our proposed AU co-occurrence constraint 3DCNN based AU recognition approach outperforms current approaches and demonstrate the effectiveness of taking use of AUs relationship in AU recognition.

The Basic Study on making mono-phone for Korean Speech Recognition (한국어 음성 인식을 위한 mono-phone 구성의 기초 연구)

  • Hwang YoungSoo;Song Minsuck
    • Proceedings of the Acoustical Society of Korea Conference
    • /
    • autumn
    • /
    • pp.45-48
    • /
    • 2000
  • In the case of making large vocabulary speech recognition system, it is better to use the segment than the syllable or the word as the recognition unit. In this paper, we study on the basis of making mono-phone for Korean speech recognition. For experiments, we use the speech toolkit of OGI in U.S.A. The result shows that the recognition rate of :he case in which the diphthong is established as a single unit is superior to that of the case in which the diphthong is established as two units, i.e. a glide plus a vowel. And also, the recognition rate by the number of consonants is a little different.

  • PDF

The Tendencies in Apartment Inhabitants' Recognition of Landscape Elements (조망 경관에 대한 아파트 거주자들의 인지 특성)

  • Lee, Sang-Bok;Moon, Ji-Won;Ha, Jae-Myung
    • Proceeding of Spring/Autumn Annual Conference of KHA
    • /
    • 2006.11a
    • /
    • pp.248-252
    • /
    • 2006
  • This study is intended to understand the intrinsic attributes of the view from the apartment unit in consideration of the diverse and complex elements of the view. To this end, the Questionnaire survey was conducted to identify the tendency in the recognition by apartment dwellers. The Questionnaire survey was conducted for the apartment residents to identify their interest in and the general trend in their recognition of the view from the living rooms of their housing unit, where Questionnaire items regarding landscape elements, the distances to and location of the landscape elements, and floor locations were compiled on the basis of the results from the field survey in the previous study. Consequently, the following results have been derived. 1) Apartment residents recognize not only natural landscape elements but also artificial elements, and prefer natural elements to artificial ones. 2) It is also indicated that they recognize the distances to and locations of landscape elements and that the satisfaction for the distance and location varies depending on the type of the landscape elements. 3) Furthermore, the floor of each unit is shown to result in certain differences in the recognized landscape elements. The cross-analysis between the floor and satisfaction indicates that the higher the floor, the more satisfied the residents are with the view.

  • PDF

A Study on Recognition Units for Korean Speech Recognition (한국어 분절음 인식을 위한 인식 단위에 대한 연구)

  • ;;Michael W. Macon
    • The Journal of the Acoustical Society of Korea
    • /
    • v.19 no.6
    • /
    • pp.47-52
    • /
    • 2000
  • In the case of making large vocabulary speech recognition system, it is better to use the segment than the syllable or the word as the recognition mit. In this paper, we study on the proper recognition units for Korean speech recognition. For experiments, we use the speech toolkit of OGI in U.S.A. The result shows that the recognition rate of the case in which the diphthong is established as a single unit is superior to that of the case in which the diphthong is established as two units, i.e. a glide plus a vowel. And also, the recognition rate of the case in which the biphone is used as the recognition unit is better than that of the case in which the mono-phoneme is used.

  • PDF