• Title/Summary/Keyword: word network analysis

Search Result 379, Processing Time 0.026 seconds

An Artificial Neural Network Based Phrase Network Construction Method for Structuring Facility Error Types (설비 오류 유형 구조화를 위한 인공신경망 기반 구절 네트워크 구축 방법)

  • Roh, Younghoon;Choi, Eunyoung;Choi, Yerim
    • Journal of Internet Computing and Services
    • /
    • v.19 no.6
    • /
    • pp.21-29
    • /
    • 2018
  • In the era of the 4-th industrial revolution, the concept of smart factory is emerging. There are efforts to predict the occurrences of facility errors which have negative effects on the utilization and productivity by using data analysis. Data composed of the situation of a facility error and the type of the error, called the facility error log, is required for the prediction. However, in many manufacturing companies, the types of facility error are not precisely defined and categorized. The worker who operates the facilities writes the type of facility error in the form with unstructured text based on his or her empirical judgement. That makes it impossible to analyze data. Therefore, this paper proposes a framework for constructing a phrase network to support the identification and classification of facility error types by using facility error logs written by operators. Specifically, phrase indicating the types are extracted from text data by using dictionary which classifies terms by their usage. Then, a phrase network is constructed by calculating the similarity between the extracted phrase. The performance of the proposed method was evaluated by using real-world facility error logs. It is expected that the proposed method will contribute to the accurate identification of error types and to the prediction of facility errors.

Trends Analysis on Research Articles of the Sharing Economy through a Meta Study Based on Big Data Analytics (빅데이터 분석 기반의 메타스터디를 통해 본 공유경제에 대한 학술연구 동향 분석)

  • Kim, Ki-youn
    • Journal of Internet Computing and Services
    • /
    • v.21 no.4
    • /
    • pp.97-107
    • /
    • 2020
  • This study aims to conduct a comprehensive meta-study from the perspective of content analysis to explore trends in Korean academic research on the sharing economy by using the big data analytics. Comprehensive meta-analysis methodology can examine the entire set of research results historically and wholly to illuminate the tendency or properties of the overall research trend. Academic research related to the sharing economy first appeared in the year in which Professor Lawrence Lessig introduced the concept of the sharing economy to the world in 2008, but research began in earnest in 2013. In particular, between 2006 and 2008, research improved dramatically. In order to grasp the overall flow of domestic academic research of trends, 8 years of papers from 2013 to the present have been selected as target analysis papers, focusing on titles, keywords, and abstracts using database of electronic journals. Big data analysis was performed in the order of cleaning, analysis, and visualization of the collected data to derive research trends and insights by year and type of literature. We used Python3.7 and Textom analysis tools for data preprocessing, text mining, and metrics frequency analysis for key word extraction, and N-gram chart, centrality and social network analysis and CONCOR clustering visualization based on UCINET6/NetDraw, Textom program, the keywords clustered into 8 groups were used to derive the typologies of each research trend. The outcomes of this study will provide useful theoretical insights and guideline to future studies.

Development and Validation of the Letter-unit based Korean Sentimental Analysis Model Using Convolution Neural Network (회선 신경망을 활용한 자모 단위 한국형 감성 분석 모델 개발 및 검증)

  • Sung, Wonkyung;An, Jaeyoung;Lee, Choong C.
    • The Journal of Society for e-Business Studies
    • /
    • v.25 no.1
    • /
    • pp.13-33
    • /
    • 2020
  • This study proposes a Korean sentimental analysis algorithm that utilizes a letter-unit embedding and convolutional neural networks. Sentimental analysis is a natural language processing technique for subjective data analysis, such as a person's attitude, opinion, and propensity, as shown in the text. Recently, Korean sentimental analysis research has been steadily increased. However, it has failed to use a general-purpose sentimental dictionary and has built-up and used its own sentimental dictionary in each field. The problem with this phenomenon is that it does not conform to the characteristics of Korean. In this study, we have developed a model for analyzing emotions by producing syllable vectors based on the onset, peak, and coda, excluding morphology analysis during the emotional analysis procedure. As a result, we were able to minimize the problem of word learning and the problem of unregistered words, and the accuracy of the model was 88%. The model is less influenced by the unstructured nature of the input data and allows for polarized classification according to the context of the text. We hope that through this developed model will be easier for non-experts who wish to perform Korean sentimental analysis.

Extracting Core Events Based on Timeline and Retweet Analysis in Twitter Corpus (트위터 문서에서 시간 및 리트윗 분석을 통한 핵심 사건 추출)

  • Tsolmon, Bayar;Lee, Kyung-Soon
    • KIPS Transactions on Software and Data Engineering
    • /
    • v.1 no.1
    • /
    • pp.69-74
    • /
    • 2012
  • Many internet users attempt to focus on the issues which have posted on social network services in a very short time. When some social big issue or event occurred, it will affect the number of comments and retweet on that day in twitter. In this paper, we propose the method of extracting core events based on timeline analysis, sentiment feature and retweet information in twitter data. To validate our method, we have compared the methods using only the frequency of words, word frequency with sentiment analysis, using only chi-square method and using sentiment analysis with chi-square method. For justification of the proposed approach, we have evaluated accuracy of correct answers in top 10 results. The proposed method achieved 94.9% performance. The experimental results show that the proposed method is effective for extracting core events in twitter corpus.

A Study on the Consumer's Perception of HiSeoul Fashion Show Using Big Data Analysis (빅데이터 분석을 활용한 하이서울패션쇼에 대한 소비자 인식 조사)

  • Han, Ki Hyang
    • Journal of Fashion Business
    • /
    • v.23 no.5
    • /
    • pp.81-95
    • /
    • 2019
  • The purpose of this study is to research consumers' perception of the HiSeoul fashion show, which is being used by new designers as a means of promotion, and to propose a strategy for revitalizing new designer brands. This was done in order to secure basic data from fashion consumers, to help guide marketing strategies and promote rising designers. In this research, the consumers' perception of HiSeoul fashion show was verified using text-mining, data refinement and word clouding that was undertaken by TEXTOM3.0. Also, semantic network analysis, CONCOR analysis and visualization of the analysis results were performed using Ucinet 6.0 and NetDraw. "HiSeoul fashion show" was used as the keyword for text-mining and data was collected from March 1, 2018 to April 30, 2019. Using frequency analysis, TF-IDF, and N-gram, it was also shown that consumers are aware of places where shows are held, such as DDP and Igansumun. It was also revealed that consumers recognize rising designer brands, designer's names, the names of guests attending the show and the photo times. This study is meaningful in that it not only confirmed consumers' interest in new designer brands participating in the HiSeoul Fashion Show through big data but also confirmed that it is available as a marketing strategy to boost brand sales. This study suggests using HiSeoul show room to induce consumer sales, or inviting guests that match the brand image to promote them on SNS on the day the show is held for a marketing strategy.

Classification and analysis of error types for deep learning-based Korean spelling correction (딥러닝 기반 한국어 맞춤법 교정을 위한 오류 유형 분류 및 분석)

  • Koo, Seonmin;Park, Chanjun;So, Aram;Lim, Heuiseok
    • Journal of the Korea Convergence Society
    • /
    • v.12 no.12
    • /
    • pp.65-74
    • /
    • 2021
  • Recently, studies on Korean spelling correction have been actively conducted based on machine translation and automatic noise generation. These methods generate noise and use as train and data set. This has limitation in that it is difficult to accurately measure performance because it is unlikely that noise other than the noise used for learning is included in the test set In addition, there is no practical error type standard, so the type of error used in each study is different, making qualitative analysis difficult. This paper proposes new 'error type classification' for deep learning-based Korean spelling correction research, and error analysis perform on existing commercialized Korean spelling correctors (System A, B, C). As a result of analysis, it was found the three correction systems did not perform well in correcting other error types presented in this paper other than spacing, and hardly recognized errors in word order or tense.

A Bibliometric Approach for Department-Level Disciplinary Analysis and Science Mapping of Research Output Using Multiple Classification Schemes

  • Gautam, Pitambar
    • Journal of Contemporary Eastern Asia
    • /
    • v.18 no.1
    • /
    • pp.7-29
    • /
    • 2019
  • This study describes an approach for comparative bibliometric analysis of scientific publications related to (i) individual or several departments comprising a university, and (ii) broader integrated subject areas using multiple disciplinary schemes. It uses a custom dataset of scientific publications (ca. 15,000 articles and reviews, published during 2009-2013, and recorded in the Web of Science Core Collections) with author affiliations to the research departments, dedicated to science, technology, engineering, mathematics, and medicine (STEMM), of a comprehensive university. The dataset was subjected, at first, to the department level and discipline level analyses using the newly available KAKEN-L3 classification (based on MEXT/JSPS Grants-in-Aid system), hierarchical clustering, correspondence analysis to decipher the major departmental and disciplinary clusters, and visualization of the department-discipline relationships using two-dimensional stacked bar diagrams. The next step involved the creation of subsets covering integrated subject areas and a comparative analysis of departmental contributions to a specific area (medical, health and life science) using several disciplinary schemes: Essential Science Indicators (ESI) 22 research fields, SCOPUS 27 subject areas, OECD Frascati 38 subordinate research fields, and KAKEN-L3 66 subject categories. To illustrate the effective use of the science mapping techniques, the same subset for medical, health and life science area was subjected to network analyses for co-occurrences of keywords, bibliographic coupling of the publication sources, and co-citation of sources in the reference lists. The science mapping approach demonstrates the ways to extract information on the prolific research themes, the most frequently used journals for publishing research findings, and the knowledge base underlying the research activities covered by the publications concerned.

The Meaning of Economic Activity of Middle-aged Men using Big Data

  • Sim, Yu Jeong;Lim, Ahn-Na
    • International journal of advanced smart convergence
    • /
    • v.9 no.3
    • /
    • pp.176-182
    • /
    • 2020
  • In this paper, to analyze the meaning of middle-aged men's economic activities, TEXTOM was used to analyze them. The data collection period is set from 2017 to 2019. Among the collected data, 100 refined words were converted into a matrix in which the degree of social connection was calculated, and the keyword network analysis was performed again with the NetDraw program. According to the study, middle-aged men put more meaning on their current work and family than their future retirement. Also, the related word commonly included in the top five for all three years was 'work'. Related words commonly included in the top 10 were 'old age', 'family', and 'work', and in 2018 and 2019, 'health' was included in the top 10. As a result of this, the middle-aged men living in the modern age are the generation who keep their families through economic activities and are increasingly interested in health and prepare for retirement. Therefore, policy support for stable economic activities is needed to improve the quality of life for middle-aged men. It is necessary to extend the retirement age, expand jobs and provide effective vocational training so that it can handle its role as the head of a family. In addition, measures should be taken to reduce the wage gap between highly skilled and low-skilled workers.

Cross-Enrichment of the Heterogenous Ontologies Through Mapping Their Conceptual Structures: the Case of Sejong Semantic Classes and KorLexNoun 1.5 (이종 개념체계의 상호보완방안 연구 - 세종의미부류와 KorLexNoun 1.5 의 사상을 중심으로)

  • Bae, Sun-Mee;Yoon, Ae-Sun
    • Language and Information
    • /
    • v.14 no.1
    • /
    • pp.165-196
    • /
    • 2010
  • The primary goal of this paper is to propose methods of enriching two heterogeneous ontologies: Sejong Semantic Classes (SJSC) and KorLexNoun 1.5 (KLN). In order to achieve this goal, this study introduces the pros and cons of two ontologies, and analyzes the error patterns found during the fine-grained manual mapping processes between them. Error patterns can be classified into four types: (1) structural defectives involved in node branching, (2) errors in assigning the semantic classes, (3) deficiency in providing linguistic information, and (4) lack of the lexical units representing specific concepts. According to these error patterns, we propose different solutions in order to correct the node branching defectives and the semantic class assignment, to complement the deficiency of linguistic information, and to increase the number of lexical units suitably allotted to their corresponding concepts. Using the results of this study, we can obtain more enriched ontologies by correcting the defects and errors in each ontology, which will lead to the enhancement of practicality for syntactic and semantic analysis.

  • PDF

Influencing Knowledge Sharing on Social Media: A Gender Perspective

  • Jae Hoon Choi;Ronald Ramirez;Dawn G. Gregg;Judy E. Scott;Kuo-Hao Lee
    • Asia pacific journal of information systems
    • /
    • v.30 no.3
    • /
    • pp.513-531
    • /
    • 2020
  • Online Word-of-Mouth communication, or eWOM, has dramatically changed the way people network, interact, and share knowledge. Studies have examined why consumers choose to share knowledge online, especially online product reviews, as well as the motivations of individuals to share product ideas online. However, the role of gender in shaping the motivation and types of knowledge shared online has been given little consideration. Using concepts from Social Exchange Theory and the Theory of Reasoned Action, we address this research gap by developing and testing a model of gender's influence on knowledge sharing in a social media context. A PLS analysis of survey data from 257 students indicates that reputation, altruism, and subjective norms are key motivators for knowledge sharing intention in social media. More importantly, that gender plays a moderating role within the motivation-knowledge sharing relationship. We also find that subjective norms have a greater impact on knowledge sharing with women than with men. Collectively, our research results highlight individualized factors for improving customer participation in external facing social media for marketing and product innovation.