• Title/Summary/Keyword: Neighbor embedding

Search Result 25, Processing Time 0.016 seconds

Decision support system for underground coal pillar stability using unsupervised and supervised machine learning approaches

  • Kamran, Muhammad;Shahani, Niaz Muhammad;Armaghani, Danial Jahed
    • Geomechanics and Engineering
    • /
    • v.30 no.2
    • /
    • pp.107-121
    • /
    • 2022
  • Coal pillar assessment is of broad importance to underground engineering structure, as the pillar failure can lead to enormous disasters. Because of the highly non-linear correlation between the pillar failure and its influential attributes, conventional forecasting techniques cannot generate accurate outcomes. To approximate the complex behavior of coal pillar, this paper elucidates a new idea to forecast the underground coal pillar stability using combined unsupervised-supervised learning. In order to build a database of the study, a total of 90 patterns of pillar cases were collected from authentic engineering structures. A state-of-the art feature depletion method, t-distribution symmetric neighbor embedding (t-SNE) has been employed to reduce significance of actual data features. Consequently, an unsupervised machine learning technique K-mean clustering was followed to reassign the t-SNE dimensionality reduced data in order to compute the relative class of coal pillar cases. Following that, the reassign dataset was divided into two parts: 70 percent for training dataset and 30 percent for testing dataset, respectively. The accuracy of the predicted data was then examined using support vector classifier (SVC) model performance measures such as precision, recall, and f1-score. As a result, the proposed model can be employed for properly predicting the pillar failure class in a variety of underground rock engineering projects.

Odorant receptors in cancer

  • Chung, Chan;Cho, Hee Jin;Lee, ChaeEun;Koo, JaeHyung
    • BMB Reports
    • /
    • v.55 no.2
    • /
    • pp.72-80
    • /
    • 2022
  • Odorant receptors (ORs), the largest subfamily of G protein-coupled receptors, detect odorants in the nose. In addition, ORs were recently shown to be expressed in many nonolfactory tissues and cells, indicating that these receptors have physiological and pathophysiological roles beyond olfaction. Many ORs are expressed by tumor cells and tissues, suggesting that they may be associated with cancer progression or may be cancer biomarkers. This review describes OR expression in various types of cancer and the association of these receptors with various types of signaling mechanisms. In addition, the clinical relevance and significance of the levels of OR expression were evaluated. Namely, levels of OR expression in cancer were analyzed based on RNA-sequencing data reported in the Cancer Genome Atlas; OR expression patterns were visualized using t-distributed stochastic neighbor embedding (t-SNE); and the associations between patient survival and levels of OR expression were analyzed. These analyses of the relationships between patient survival and expression patterns obtained from an open mRNA database in cancer patients indicate that ORs may be cancer biomarkers and therapeutic targets.

Stochastic Strength Analysis according to Initial Void Defects in Composite Materials (복합재 초기 공극 결함에 따른 횡하중 강도 확률론적 분석)

  • Seung-Min Ji;Sung-Wook Cho;S.S. Cheon
    • Composites Research
    • /
    • v.37 no.3
    • /
    • pp.179-185
    • /
    • 2024
  • This study quantitatively evaluated and investigated the changes in transverse tensile strength of unidirectional fiber-reinforced composites with initial void defects using a Representative Volume Element (RVE) model. After calculating the appropriate sample size based on margin of error and confidence level for initial void defects, a sample group of 5000 RVE models with initial void defects was generated. Dimensional reduction and density-based clustering analysis were conducted on the sample group to assess similarity, confirming and verifying that the sample group was unbiased. The validated sample analysis results were represented using a Weibull distribution, allowing them to be applied to the reliability analysis of composite structures.

Accuracy of one-step automated orthodontic diagnosis model using a convolutional neural network and lateral cephalogram images with different qualities obtained from nationwide multi-hospitals

  • Yim, Sunjin;Kim, Sungchul;Kim, Inhwan;Park, Jae-Woo;Cho, Jin-Hyoung;Hong, Mihee;Kang, Kyung-Hwa;Kim, Minji;Kim, Su-Jung;Kim, Yoon-Ji;Kim, Young Ho;Lim, Sung-Hoon;Sung, Sang Jin;Kim, Namkug;Baek, Seung-Hak
    • The korean journal of orthodontics
    • /
    • v.52 no.1
    • /
    • pp.3-19
    • /
    • 2022
  • Objective: The purpose of this study was to investigate the accuracy of one-step automated orthodontic diagnosis of skeletodental discrepancies using a convolutional neural network (CNN) and lateral cephalogram images with different qualities from nationwide multi-hospitals. Methods: Among 2,174 lateral cephalograms, 1,993 cephalograms from two hospitals were used for training and internal test sets and 181 cephalograms from eight other hospitals were used for an external test set. They were divided into three classification groups according to anteroposterior skeletal discrepancies (Class I, II, and III), vertical skeletal discrepancies (normodivergent, hypodivergent, and hyperdivergent patterns), and vertical dental discrepancies (normal overbite, deep bite, and open bite) as a gold standard. Pre-trained DenseNet-169 was used as a CNN classifier model. Diagnostic performance was evaluated by receiver operating characteristic (ROC) analysis, t-stochastic neighbor embedding (t-SNE), and gradient-weighted class activation mapping (Grad-CAM). Results: In the ROC analysis, the mean area under the curve and the mean accuracy of all classifications were high with both internal and external test sets (all, > 0.89 and > 0.80). In the t-SNE analysis, our model succeeded in creating good separation between three classification groups. Grad-CAM figures showed differences in the location and size of the focus areas between three classification groups in each diagnosis. Conclusions: Since the accuracy of our model was validated with both internal and external test sets, it shows the possible usefulness of a one-step automated orthodontic diagnosis tool using a CNN model. However, it still needs technical improvement in terms of classifying vertical dental discrepancies.

Research Trends in Record Management Using Unstructured Text Data Analysis (비정형 텍스트 데이터 분석을 활용한 기록관리 분야 연구동향)

  • Deokyong Hong;Junseok Heo
    • Journal of Korean Society of Archives and Records Management
    • /
    • v.23 no.4
    • /
    • pp.73-89
    • /
    • 2023
  • This study aims to analyze the frequency of keywords used in Korean abstracts, which are unstructured text data in the domestic record management research field, using text mining techniques to identify domestic record management research trends through distance analysis between keywords. To this end, 1,157 keywords of 77,578 journals were visualized by extracting 1,157 articles from 7 journal types (28 types) searched by major category (complex study) and middle category (literature informatics) from the institutional statistics (registered site, candidate site) of the Korean Citation Index (KCI). Analysis of t-Distributed Stochastic Neighbor Embedding (t-SNE) and Scattertext using Word2vec was performed. As a result of the analysis, first, it was confirmed that keywords such as "record management" (889 times), "analysis" (888 times), "archive" (742 times), "record" (562 times), and "utilization" (449 times) were treated as significant topics by researchers. Second, Word2vec analysis generated vector representations between keywords, and similarity distances were investigated and visualized using t-SNE and Scattertext. In the visualization results, the research area for record management was divided into two groups, with keywords such as "archiving," "national record management," "standardization," "official documents," and "record management systems" occurring frequently in the first group (past). On the other hand, keywords such as "community," "data," "record information service," "online," and "digital archives" in the second group (current) were garnering substantial focus.