통합 검색 | Korea Science

Optimization of Domain-Independent Classification Framework for Mood Classification

Choi, Sung-Pil;Jung, Yu-Chul;Myaeng, Sung-Hyon
- Journal of Information Processing Systems
- /
- 제3권2호
- /
- pp.73-81
- /
- 2007
In this paper, we introduce a domain-independent classification framework based on both k-nearest neighbor and Naive Bayesian classification algorithms. The architecture of our system is simple and modularized in that each sub-module of the system could be changed or improved efficiently. Moreover, it provides various feature selection mechanisms to be applied to optimize the general-purpose classifiers for a specific domain. As for the enhanced classification performance, our system provides conditional probability boosting (CPB) mechanism which could be used in various domains. In the mood classification domain, our optimized framework using the CPB algorithm showed 1% of improvement in precision and 2% in recall compared with the baseline.
https://doi.org/10.3745/JIPS.2008.3.2.073 인용 PDF KSCI

High Accuracy Classification Methods for Multi-Temporal Images

Hong, Sun Pyo;Jeon, Dong Keun
- The Journal of the Acoustical Society of Korea
- /
- 제16권1E호
- /
- pp.3-8
- /
- 1997
Three new classification methods for multi temporal images are proposed. They are named as a likelihood addition method, a likelihood majority method and a Dempster-Shafer's rule method. Basic strategies using these methods are to calculate likelihoods for each temporal data and to combine obtained likelihoods for final classification. These three methods use different combining algorithms. From classification experiments, following results were obtained. The method based on Dempster-Shafer's rule of combination showed about 12% improvement of classification accuracies compared to a conventional method. This method needed about 16% more processing times than that of a conventional method. The other two proposed method showed 1% to 5% increase of classification accuracies. However processing times of these two proposed method showed 1% to 5% increase of classification accuracies. However processing times of these two methods are almost the same with that of a conventional method. Among the newly proposed three methods, the Dempster-Shafer's rule method showed the highest classification accuracies with more processing time than those of other methods.
PDF

An Improved Text Classification Method for Sentiment Classification

Wang, Guangxing;Shin, Seong Yoon
- Journal of information and communication convergence engineering
- /
- 제17권1호
- /
- pp.41-48
- /
- 2019
In recent years, sentiment analysis research has become popular. The research results of sentiment analysis have achieved remarkable results in practical applications, such as in Amazon's book recommendation system and the North American movie box office evaluation system. Analyzing big data based on user preferences and evaluations and recommending hot-selling books and hot-rated movies to users in a targeted manner greatly improve book sales and attendance rate in movies [1, 2]. However, traditional machine learning-based sentiment analysis methods such as the Classification and Regression Tree (CART), Support Vector Machine (SVM), and k-nearest neighbor classification (kNN) had performed poorly in accuracy. In this paper, an improved kNN classification method is proposed. Through the improved method and normalizing of data, the purpose of improving accuracy is achieved. Subsequently, the three classification algorithms and the improved algorithm were compared based on experimental data. Experiments show that the improved method performs best in the kNN classification method, with an accuracy rate of 11.5% and a precision rate of 20.3%.
https://doi.org/10.6109/jicce.2019.17.1.41 인용 PDF KSCI HTML

텍스트 분류 기법의 발전 (Enhancement of Text Classification Method)

신광성;신성윤
- 한국정보통신학회:학술대회논문집
- /
- 한국정보통신학회 2019년도 춘계학술대회
- /
- pp.155-156
- /
- 2019
Classification and Regression Tree (CART), SVM (Support Vector Machine) 및 k-nearest neighbor classification (kNN)과 같은 기존 기계 학습 기반 감정 분석 방법은 정확성이 떨어졌습니다. 본 논문에서는 개선 된 kNN 분류 방법을 제안한다. 개선 된 방법 및 데이터 정규화를 통해 정확성 향상의 목적이 달성됩니다. 그 후, 3 가지 분류 알고리즘과 개선 된 알고리즘을 실험 데이터에 기초하여 비교 하였다.
PDF

Intelligent System for the Prediction of Heart Diseases Using Machine Learning Algorithms with Anew Mixed Feature Creation (MFC) technique

Rawia Elarabi;Abdelrahman Elsharif Karrar;Murtada El-mukashfi El-taher
- International Journal of Computer Science & Network Security
- /
- 제23권5호
- /
- pp.148-162
- /
- 2023
Classification systems can significantly assist the medical sector by allowing for the precise and quick diagnosis of diseases. As a result, both doctors and patients will save time. A possible way for identifying risk variables is to use machine learning algorithms. Non-surgical technologies, such as machine learning, are trustworthy and effective in categorizing healthy and heart-disease patients, and they save time and effort. The goal of this study is to create a medical intelligent decision support system based on machine learning for the diagnosis of heart disease. We have used a mixed feature creation (MFC) technique to generate new features from the UCI Cleveland Cardiology dataset. We select the most suitable features by using Least Absolute Shrinkage and Selection Operator (LASSO), Recursive Feature Elimination with Random Forest feature selection (RFE-RF) and the best features of both LASSO RFE-RF (BLR) techniques. Cross-validated and grid-search methods are used to optimize the parameters of the estimator used in applying these algorithms. and classifier performance assessment metrics including classification accuracy, specificity, sensitivity, precision, and F1-Score, of each classification model, along with execution time and RMSE the results are presented independently for comparison. Our proposed work finds the best potential outcome across all available prediction models and improves the system's performance, allowing physicians to diagnose heart patients more accurately.
https://doi.org/10.22937/IJCSNS.2023.23.5.17 인용 PDF

평행사변형 분류 알고리즘의 성능에 대한 연구 (A Study on the Performance of Parallelepiped Classification Algorithm)

용환기
- 한국지리정보학회지
- /
- 제4권4호
- /
- pp.1-7
- /
- 2001
위성영상은 GIS 정보획득을 위한 가장 중요한 초기자료로서, 이로부터 주제도와 같은 유용한 정보를 추출하기 위해서는 위성영상 즉 다중스펙트럼 영상을 목적에 적합하게 분류하는 처리과정이 필요하다. 위성영상의 분류기법은 크게 감독기법과 무감독기법으로 나뉘는데, 본 논문에서는 감독분류기법 중의 하나인 평행사변형 알고리즘에서 군집의 초기값 설정이 알고리즘의 성능에 미치는 영향을 분석한다. 본 연구에서는 우선 직렬컴퓨터에서 평행사변형 알고리즘의 성능과 초기값 변화와의 관계를 살펴보고, 이를 확장하여 MIMD 병렬구조 컴퓨터 모델을 사용한 경우에 초기값의 변화가 평행사변형 알고리즘의 성능에 미치는 영향을 분석한다. 평행사변형 알고리즘의 성능은 초기값의 설정에 따라 직렬구조의 컴퓨터를 사용하는 경우에는 최고 2.4배, 그리고 MIMD 병렬구조 모델을 사용한 경우에는 최고 2.5배의 성능 향상을 보였다. 전산모의실험을 통해 위성영상의 감독분류기법에서 초기값이 평행사변형 분류알고리즘의 성능에 상당한 영향을 미치며, 직렬컴퓨터와 MIMD 병렬컴퓨터에서 초기값의 적절한 설정을 통해 분류기법의 성능이 향상됨을 확인하였다.
PDF

분산커널 기반의 퍼지 c-평균을 이용한 음악 데이터의 장르 분류 (Classification of Music Data using Fuzzy c-Means with Divergence Kernel)

박동철
- 전자공학회논문지CI
- /
- 제46권3호
- /
- pp.1-7
- /
- 2009
본 논문은 효율적인 음악 데이터의 분류를 위한 방법으로 분산커널 기반의 퍼지 c-평균을 이용한 분류기 모델을 제안한다. 분산 커널 기반의 퍼지 c-평균은 주어진 오디오 데이터에서 추출된 특징벡터의 평균과 공분산 정보를 동시에 이용하여 기존의 평균값만을 사용하는 방식에 비해 성능을 월등히 향상시킬 수 있는 장점이 있다. 사용된 방식은 확률적 분포로 주어지는 데이터 사이의 거리를 분산거리척도로 측정하고, 복잡한 분류 경계를 단순화 시키는데 효율적인 커널 개념을 사용함으로서 분류의 정확도를 극대화 시킬 수 있는 장점이 있다. 제안하는 분류기의 성능을 평가하기 위하여 고전음악, 컨트리음악, 힙합, 재즈의 4개의 장르 음악데이터를 총 1200개 수집하여 실험을 진행하였다. 실험의 결과 제안된 분산커널 기반의 퍼지 c-평균을 이용하는 분류기는 기존의 방식과 비교하여 분류정확도에서 평균적으로 17.73%-21.84%의 성능향상을 보여준다.
PDF KSCI

국내 학술논문 주제 분류 알고리즘 비교 및 분석 (Comparison and Analysis of Subject Classification for Domestic Research Data)

최원준;설재욱;정희석;윤화묵
- 한국콘텐츠학회논문지
- /
- 제18권8호
- /
- pp.178-186
- /
- 2018
학술정보 성과물을 서비스하기 위하여 논문 단위의 주제 분류는 필수가 된다. 하지만 현재까지 저널 단위의 주제 분류가 되어 있으며 기사 단위의 주제 분류가 서비스되는 곳은 많지 않다. 국내 성과물 중에서 학술 논문의 경우 주제 분류가 있으면 좀 더 큰 영역의 서비스를 담당할 수 있고 범위를 정해서 서비스 할 수 있기 때문에 무엇보다 중요한 정보가 된다. 하지만, 분야 별 주제를 분류하는 문제는 다양한 분야의 전문가의 손이 필요하고 정확도를 높이기 위해서 다양한 방법의 검증이 필요하다. 본 논문에서는 정답이 알려져 있지 않은 상태에서의 정답을 찾는 비지도 학습 알고리즘을 활용해서 주제 분류를 시도해 보고 연관도와 복잡도를 활용해서 주제 분류 알고리즘의 결과를 비교해 보고자 한다. 비지도 학습 알고리즘은 주제 분류 방법으로 잘 알려진 Hierarchical Dirichlet Precess(HDP). Latent Dirichlet Allocation(LDA), Latent Semantic Indexing(LSI) 알고리즘을 활용하여 성능을 분석해 보았다.
https://doi.org/10.5392/JKCA.2018.18.08.178 인용 PDF KSCI

패킷 분류를 위한 스마트 셋-프루닝 트라이 (A Smart Set-Pruning Trie for Packet Classification)

민세원;이나라;임혜숙
- 한국통신학회논문지
- /
- 제36권11B호
- /
- pp.1285-1296
- /
- 2011
패킷분류는 라우터의 가장 기본적이면서도 중요한 기능 중의 하나이며, 실시간 전송을 요구하는 새로운 인터넷 응용 프로그램의 등장과 더불어 그 중요성이 더욱 커지고 있다. 패킷분류는 입력 패킷에 대하여 선속도로 이루어져야 하며, 여러 헤더 필드에 대해 다차원 검색을 수행해야 하기 때문에 라우터 설계의 어려운 문제 중에 하나이다. 고속의 패킷분류를 제공하기 위한 다양한 패킷분류 알고리즘이 제안되어 왔으며, 그 중 계층적 접근 방식을 사용한 알고리즘은 하나의 필드에 대하여 검색이 수행될 때마다 많은 검색 영역이 제거되기 때문에 효율적이다. 그러나 계층적 구조는 역추적이라는 문제를 내재하고 있으며, 이를 해결하기 위해 사용되는 셋-프루닝 트라이나그리드-오브-트라이는 지나치게 많은 노드 복사를 야기하거나, 선-계산이라는 복잡한 과정을 요구한다. 본 논문에서는 셋-프루닝 하위 트라이의 간단한 합병을 통하여 복사되는 노드의 개수를 줄일 수 있는 스마트 셋-프루닝 구조를 제안한다. 시뮬레이션 결과 제안된 구조는 셋-프루닝 트라이와 비교하여 복사되는 노드 수 및 룰 수가 2-8% 줄어듦을 확인하였다.
https://doi.org/10.7840/KICS.2011.36B.11.1285 인용 PDF KSCI

인터넷 라우터에서의 패킷 분류를 위한 2차원 이진 검색 트리 (Two-dimensional Binary Search Tree for Packet Classification at Internet Routers)

이고은;임혜숙
- 전자공학회논문지
- /
- 제52권6호
- /
- pp.21-31
- /
- 2015
현재의 인터넷 사용자들은 실시간으로 다양한 멀티미디어 서비스를 제공 받길 원한다. 이에 네트워크 트래픽의 속도는 매우 빨라지고 있으며, 처리하여야 하는 데이터의 양은 해마다 기하급수적으로 증가하고 있다. 데이터는 '패킷'이라는 단위의 데이터 형식으로 전송되며, 패킷분류는 인터넷 라우터의 가장 어려운 기능 중 하나로 모든 패킷에 대하여 선속도로 처리되어야 한다. 다양한 패킷 분류 알고리즘 중, 영역분할 패킷분류 알고리즘은 5개의 패킷 헤더 필드 정보를 동시에 검색할 수 있는 효율적인 알고리즘이다. 영역 분할 사분 트라이는 가장 대표적인 영역분할 패킷분류 알고리즘으로 메모리 요구량이 적은 알고리즘이지 만, 빠른 검색성능을 보장하지 못하는 단점이 있다. 본 논문에서는, 영역 분할 사분 트라이의 단점을 이진 검색 트리를 사용해 보완하는 새로운 알고리즘을 제안한다. 실험을 통하여 제안하는 알고리즘은 입력과 비교되는 룰의 수에 있어 영역 분할 사분 트라이 보다 검색 성능이 향상됨을 보았다.
https://doi.org/10.5573/ieie.2015.52.6.021 인용 PDF KSCI

검색결과 1,190건 처리시간 0.024초

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

자세히 찾기

이미지 검색 (β)