• Title/Summary/Keyword: 용어추출

Search Result 365, Processing Time 0.026 seconds

A Study on the Effects of Search Language on Web Searching Behavior: Focused on the Differences of Web Searching Pattern (검색 언어가 웹 정보검색행위에 미치는 영향에 관한 연구 - 웹 정보검색행위의 양상 차이를 중심으로 -)

  • Byun, Jeayeon
    • Journal of the Korean Society for Library and Information Science
    • /
    • v.52 no.3
    • /
    • pp.289-334
    • /
    • 2018
  • Even though information in many languages other than English is quickly increasing, English is still playing the role of the lingua franca and being accounted for the largest proportion on the web. Therefore, it is necessary to investigate the key features and differences between "information searching behavior using mother tongue as a search language" and "information searching behavior using English as a search language" of users who are non-mother tongue speakers of English to acquire more diverse and abundant information. This study conducted the experiment on the web searching which is applied in concurrent think-aloud method to examine the information searching behavior and the cognitive process in Korean search and English search through the twenty-four undergraduate students at a private university in South Korea. Based on the qualitative data, this study applied the frequency analysis to web search pattern under search language. As a result, it is active, aggressive and independent information searching behavior in Korean search, while information searching behavior in English search is passive, submissive and dependent. In Korean search, the main features are the query formulation by extract and combine the terms from various sources such as users, tasks and system, the search range adjustment in diverse level, the smooth filtering of the item selection in search engine results pages, the exploration and comparison of many items and the browsing of the overall contents of web pages. Whereas, in English search, the main features are the query formulation by the terms principally extracted from task, the search range adjustment in limitative level, the item selection by rely on the relevance between the items such as categories or links, the repetitive exploring on same item, the browsing of partial contents of web pages and the frequent use of language support tools like dictionaries or translators.

An Exploratory Study of Image Retrieval Using Aesthetic Impressions (심미적 인상을 이용한 이미지 검색에 관한 실험적 연구)

  • Yu, So-Young;Moon, Sung-Been
    • Journal of the Korean Society for information Management
    • /
    • v.21 no.4 s.54
    • /
    • pp.187-208
    • /
    • 2004
  • In this study, aesthetic impressions were used for a high-level feature of image retrieval. The term, 'aesthetic' has been studied in psychology, art, and literature. It means unconscious, instantaneous parts of visual perception and emotion. The literatures related to aesthetic impressions were reviewed and four kinds of aesthetic impressions were defined operationally : strong impression, soft impression, courteous impression, and refined impression. 66 image files of paintings were sampled randomly from 1100 paintings and low-level color features were extracted from them by a using perceptual color model(Lai, & Tait, 1998). The high-level features of an image, that is, four kinds of aesthetic impressions of each painting were measured by 4 subjects and averaged. In CBIR, 2 subjects performed image retrievals using example queries. They were asked to retrieve images by using the aesthetic impressions or the keywords. In evaluations, subjects showed that they were satisfied with the aesthetic impression-based image retrieval system on the average. And R-precision of the image retrieval with both color features and aesthetic impressions was higher than that of the image retrieval with color features only. But further studies with larger test collections and query sets should be followed for generalization of the result of this study.

An Analysis of High School Students' Analogy Generating Processes Using Think-Aloud Method (발성사고법을 활용한 고등학생의 비유 생성 과정 분석)

  • Kim, Minhwan;Kwon, Hyeoksoon;Lee, Donghwi;Noh, Taehee
    • Journal of The Korean Association For Science Education
    • /
    • v.38 no.1
    • /
    • pp.43-55
    • /
    • 2018
  • In this study, we investigated high school students' analogy generating processes using the think-aloud method. Twelve high school students in Seoul participated in this study. The students were asked to generate analogies on ionic bonding and were also interviewed after their activities. Their activities and interviews were recorded and videotaped. After classifying the analogy generating processes into the three stages-encoding, exploring sources, and mapping, several process components were identified. The analyses of the results indicated that they checked the target concept given and selected one for a salient attribute among many attributes of the target concept at the stage of encoding. After selecting the salient attribute, they translated the salient attribute that is a scientific term into an everyday term, which is named as 'extracting salient similarities.' At the stage of exploring sources, they chose the sources based on salient similarities and chose the final source through circular processes, which included the process components of 'evaluating the sources' and 'discarding the sources.' At the final stage, they added the attributes to analogs and mapping them to the attributes of the target concept, which is named as 'mapping shared attributes.' There were some cases that 'mapping shared attributes' appeared after they specified the situation of analogs or assumed new situation, which is named as 'specifying the situations.' Some students recognized unshared attributes in their analogs.

Tactile Sensibility Factors of Traditional Silk Fabrics (전통 견직물의 촉각적 감성요인)

  • Yi, Eun-Jou
    • Science of Emotion and Sensibility
    • /
    • v.10 no.1
    • /
    • pp.99-111
    • /
    • 2007
  • In order to identify tactile sensibility factors of traditional silk fabrics and to provide prediction models for the sensibility factors by mechanical properties, seventeen different traditional silk fabrics were evaluated in terms of both tactile sensation and sensibility by using a modified magnitude estimation line scale Gongdan and Newttong with lower values for surface roughness(SMD), bending rigidity(B), and compression resilience(RC) were rated as softer, smoother, fluffier, and more pliable in tactile sensation than any other traditional silk fabrics whereas Nobangju haying higher B, SMD, and tensile resilience(RT) was touched as crisper, more rustling, and springier. Three different tactile sensibility factors including 'Feminine', 'Natural', and 'Casual' were obtained significantly by grouping fifteen different tactile sensibility descriptors. In the prediction models sensibility 'Feminine' was explained positively by SMD, which was supported by the fact that both Gongdan and Newtton were perceived as more feminine. Sensibility 'Natural' that was felt stronger as for Myoungju and Sa was predicted negatively by both fabric thickness(T) and RT. Finally, RC, elongation at maximum load (EM), and T predicted sensibility 'Casual' negatively, which results in its higher factor scores for Myoungju and Shantung, respectively.

  • PDF

A Sentence Theme Allocation Scheme based on Head Driven Patterns in Encyclopedia Domain (백과사전 영역에서 중심어주도패턴에 기반한 문장주제 할당 기법)

  • Kang Bo-Young;Myaeng Sung-Hyon
    • Journal of KIISE:Software and Applications
    • /
    • v.32 no.5
    • /
    • pp.396-405
    • /
    • 2005
  • Since sentences are the basic propositional units of text, their themes would be helpful for various tasks that require knowledge about the semantic content of text. Despite the importance of determining the theme of a sentence, however, few studies have investigated the problem of automatically assigning the theme to a sentence. Therefore, we propose a sentence theme allocation scheme based on the head-driven patterns of sentences in encyclopedia. In a serious of experiments using Dusan Dong-A encyclopedia, the proposed method outperformed the baseline of the theme allocation performance. The head-driven pattern 4, which is reconfigured based on the predicate, showed superior performance in the theme allocation with the average F-score of $98.96\%$ for the training data, and $88.57\%$ for the test data.

A Study on Construction and Management Tools for Biological Named Entity Dictionary (생물학적 개체명 사전을 위한 구축 및 관리 도구에 관한 연구)

  • Jang, Hyun-Chul;Kim, Tae-Hyun;Lee, Hyun-Sook;Park, Soo-Jun;Park, Seon-Hee
    • Annual Conference of KIPS
    • /
    • 2003.11b
    • /
    • pp.853-856
    • /
    • 2003
  • 바이오 텍스트 마이닝을 위한 정보 추출의 첫 단계는 생물학적 문헌으로부터의 유전자, 단백질, 세포조직 등과 같은 생물학적 개체명의 인식이다. 생물학적 개체명의 명명법상 특징이 매우 다양하고 저자의 개성에 의해 쉽게 좌우되어 단순히 규칙이나 학습 방법 만으로는 쉽게 개체명들을 인식할 수 없다. 또한, 생물학 관련 문헌에 나오는 가능한 모든 개체명과 이들의 모든 변형을 수록하는 것은 현실적으로 불가능하므로 이를 해결하기 위해 이미 알려진 개체명에 대해서 기본적으로 사전을 탐색하고 알려지지 않은 용어들을 규칙과 통계 기반 방법을 통하여 인식하는 것이 효과적이다. 그러나 만족할 만한 수준의 양질의 사전을 구축하는 것은 쉽지 않을 뿐만 아니라 많은 비용이 소요되며, 어느 순간 만족할 만한 성능을 낼 수 있는 사전을 구축했다. 할지라도 유지 관리 하는 것이 결코 쉬운 일이 아니며 마찬가지로 많은 비용을 필요로 하게 된다. 따라서, 잘 구축된 자원으로부터 필요한 정보를 추출하여 적절한 사전을 자동으로 구축하여 활용하는 방법을 사용할 경우, 사전 구축 및 관리에 드는 많은 비용을 줄이면서도 상당히 효과적인 성능을 얻을 수 있을 것이다. 본 연구에서는 바이오 텍스트 마이닝 엔진을 위한 생물학적 개체명 사전을 자동으로 구축하고 이를 쉽게 관리하도록 하는 도구를 개발하였다.

  • PDF

User Interaction-based Graph Query Formulation and Processing (사용자 상호작용에 기반한 그래프질의 생성 및 처리)

  • Jung, Sung-Jae;Kim, Taehong;Lee, Seungwoo;Lee, Hwasik;Jung, Hanmin
    • Journal of KIISE:Databases
    • /
    • v.41 no.4
    • /
    • pp.242-248
    • /
    • 2014
  • With the rapidly growing amount of information represented in RDF format, efficient querying of RDF graph has become a fundamental challenge. SPARQL is one of the most widely used query languages for retrieving information from RDF dataset. SPARQL is not only simple in its syntax but also powerful in representation of graph pattern queries. However, users need to make a lot of efforts to understand the ontology schema of a dataset in order to compose a relevant SPARQL query. In this paper, we propose a graph query formulation and processing scheme based on ontology schema information which can be obtained by summarizing RDF graph. In the context of the proposed querying scheme, a user can interactively formulate the graph queries on the graphic user interface without making efforts to understand the ontology schema and even without learning SPARQL syntax. The graph query formulated by a user is transformed into a set of class paths, which are stored in a relational database and used as the constraint for search space reduction when the relational database executes the graph search operation. By executing the LUBM query 2, 8, and 9 over LUBM (10,0), it is shown that the proposed querying scheme returns the complete result set.

Robust Gait Recognition for Directional Variation Using Canonical View Synthesis (고유시점 재구성을 이용한 방향 변화에 강인한 게이트 인식)

  • 정승도;최병욱
    • Journal of the Institute of Electronics Engineers of Korea CI
    • /
    • v.41 no.5
    • /
    • pp.59-67
    • /
    • 2004
  • Gait is defined as a manner or characteristics of walking. Recently, the study on extracting features of the gait to identify the individual has been progressed actively, within the computer vision community. Even if the camera is fixed, gait features extracted from images are varied according to the direction of walking. In this paper, we propose the method which compensates for the drawback of the gait recognition which is dependant on the direction. First, we search a direction of walking and estimate the planar homography with simple operations. Through synthesizing canonical viewed images by using the estimated homography, viewpoint variation by the direction of walking is compensated. In this paper, we segment gait silhouette into sub-regions and use averaged feature and its variation of each region to recognition experiment. Experimental results show that the proposed method is robust for directional variation of the gait.

Method Customizing From Web-based English-Korean MT System To English-Korean MT System for Patent Documents (웹 영한 번역기로부터 특허 영한 번역기로의 특화 방법)

  • Choi, Sung-Kwon;Kwon, Oh-Woog;Lee, Ki-Young;Roh, Yoon-Hyung;Park, Sang-Kyu
    • Annual Conference on Human and Language Technology
    • /
    • 2006.10e
    • /
    • pp.57-64
    • /
    • 2006
  • 본 논문에서는 웹과 같은 일반적인 도메인의 영한 자동 번역기를 특허용 영한 자동번역기로 특화하는 방법에 대해 기술한다. 특허용 영한 파동번역기로의 특화는 다음과 같은 절차에 의해 이루어진다: 1) 대용량 특허 문서에 대한 언어학적 특성 분석, 2) 대용량 특허문서 대상 전문용어 추출 및 대역어 구축, 3) 기존 번역사전 대역어의 특화, 4) 특허문서 고유의 번역 패턴 추출 및 구축, 5) 언어학적 특성 분석에 따른 번역 엔진 모듈의 특화 및 개선, 6) 특화된 번역 지식 및 번역 엔진 모듈에 따른 번역률 평가. 이와 같은 절차에 의해 만들어진 특허 영한 자동 번역기는 특허 전문번역가의 평가에 의해 전분야 평균 81.03%의 번역률을 내었으며, 분야별로는 기계분야(80.54%), 전기전자분야(81.58%), 화학일반분야(79.92%), 의료위생분야(80.79%), 컴퓨터분야(82.29%)의 성능을 보였으며 계속 개선 중에 있다. 현재 본 논문에서 기술된 영한 특허 자동번역 시스템은 산업자원부의 특허지원센터에서 변리사 및 특허 심사관이 영어 전기전자분야 특허 문서를 검색할 때 한국어 번역서비스를 제공받도록 이용되고 있으며($\underline{http://www.ipac.or.kr}$), 2007년에는 전분야 특허문서에 대한 영한 자동번역 서비스를 제공할 예정이다.

  • PDF

Design and Implementation of Minutes Summary System Based on Word Frequency and Similarity Analysis (단어 빈도와 유사도 분석 기반의 회의록 요약 시스템 설계 및 구현)

  • Heo, Kanhgo;Yang, Jinwoo;Kim, Donghyun;Bok, Kyoungsoo;Yoo, Jaesoo
    • The Journal of the Korea Contents Association
    • /
    • v.19 no.10
    • /
    • pp.620-629
    • /
    • 2019
  • An automated minutes summary system is required to objectively summarize and classify the contents of discussions or discussions for decision making. This paper designs and implements a minutes summary system using word2vec model to complement the existing minutes summary system. The proposed system is further implemented with word2vec model to remove index words during morpheme analysis and to extract representative sentences with common opinions from documents. The proposed system automatically classifies documents collected during the meeting process and extracts representative sentences representing the agenda among various opinions. The conference host can quickly identify and manage all the agendas discussed at the meeting through the proposal system. The proposed system analyzes various agendas of large-scale debates or discussions and summarizes sentences that can be representative opinions to support fast and accurate decision making.