통합 검색 | Korea Science

Stock News Dataset Quality Assessment by Evaluating the Data Distribution and the Sentiment Prediction

Alasmari, Eman;Hamdy, Mohamed;Alyoubi, Khaled H.;Alotaibi, Fahd Saleh
- International Journal of Computer Science & Network Security
- /
- 제22권2호
- /
- pp.1-8
- /
- 2022
This work provides a reliable and classified stocks dataset merged with Saudi stock news. This dataset allows researchers to analyze and better understand the realities, impacts, and relationships between stock news and stock fluctuations. The data were collected from the Saudi stock market via the Corporate News (CN) and Historical Data Stocks (HDS) datasets. As their names suggest, CN contains news, and HDS provides information concerning how stock values change over time. Both datasets cover the period from 2011 to 2019, have 30,098 rows, and have 16 variables-four of which they share and 12 of which differ. Therefore, the combined dataset presented here includes 30,098 published news pieces and information about stock fluctuations across nine years. Stock news polarity has been interpreted in various ways by native Arabic speakers associated with the stock domain. Therefore, this polarity was categorized manually based on Arabic semantics. As the Saudi stock market massively contributes to the international economy, this dataset is essential for stock investors and analyzers. The dataset has been prepared for educational and scientific purposes, motivated by the scarcity of data describing the impact of Saudi stock news on stock activities. It will, therefore, be useful across many sectors, including stock market analytics, data mining, statistics, machine learning, and deep learning. The data evaluation is applied by testing the data distribution of the categories and the sentiment prediction-the data distribution over classes and sentiment prediction accuracy. The results show that the data distribution of the polarity over sectors is considered a balanced distribution. The NB model is developed to evaluate the data quality based on sentiment classification, proving the data reliability by achieving 68% accuracy. So, the data evaluation results ensure dataset reliability, readiness, and high quality for any usage.
https://doi.org/10.22937/IJCSNS.2022.22.2.1 인용 PDF KSCI

Forecasting LNG Freight rate with Artificial Neural Networks

Lim, Sangseop;Ahn, Young-Joong
- 한국컴퓨터정보학회논문지
- /
- 제27권7호
- /
- pp.187-194
- /
- 2022
LNG는 미래 친환경으로 가는 과도기적 에너지원으로서, 세계적인 친환경 규제, COVID-19 팬데믹, 러시아-우크라이나 전쟁 등을 계기로 엄청난 시장의 주목을 받고 있으며, 미국과 호주 등 새로운 LNG 공급처도 다양화되고 있어 LNG 스팟시장이 갈수록 커질 것으로 예상된다. 이에 반해 LNG 운송시장에 관한 연구는 그동안 소외됐었다. 본 연구는 LNG 160K 스팟운임의 단기예측에 연구를 시도하였으며 인공신경망과 ARIMA 모형을 활용하여 예측성능을 비교하였다. 본 논문의 결과, ARIMA와 인공신경망의 예측성능에 관한 우열을 가리기는 어려웠으나 ARIMA모형이 가지는 데이터 제약이 있으므로 ANN의 상대적인 자유로운 제약조건을 고려하면 LNG 160K 스팟운임 예측에 활용 가능성을 확인하였다. 본 논문은 LNG 160K 스팟운임에 관하여 인공신경망을 적용한 최초의 시도로서 학문적인 의의가 있으며, 스팟운임의 단기예측 정확성을 높여 시장 참여자들의 단기투자 의사결정의 질을 높일 수 있다는 측면에서 실무적인 기여를 할 수 있을 것으로 기대된다.
https://doi.org/10.9708/jksci.2022.27.07.187 인용 PDF KSCI HTML

A Study on the Prediction Model for International Trade Payment Using Logistic Regression

Joo, Hye-Young;Lee, Dong-Jun
- Journal of Korea Trade
- /
- 제25권2호
- /
- pp.111-133
- /
- 2021
Purpose - Although remittance payment in international trade settlements has played a bigger role in recent years, scant research is being done. This study is to zero in on analyzing determinants of international trade payments focused on remittance by constructing a payment prediction model. Design/methodology - This study categorizes the types of trade payments into advance remittance, post remittance, linked remittance, letter of credit, and mixed payment, and analyzes these after constructing a logit model. For empirical analysis, 147 survey data were collected for export manufacturers in Korea, and binominal logistic regression analysis was used to analyze the type of payment method the exporter chooses for trade transactions. Findings - The likelihood of choosing advance remittance increased as the exporters had non-recovery experiences with payments, and decreased as the market power of importers increased. The possibility of post remittance increased when the export amount was large and the character of the buyer was reliable. In the case of linked remittance, it was highly likely to be selected when payment efficiency was important in trade settlement. In addition, when competition among companies in the global market is intense and market uncertainty is high, the possibility of using a letter of credit decreases. It was also found that the greater the export amount, the greater the possibility of choosing advance remittance, and even if the transaction period was longer, exporters using a letter of credit continued to use it. Originality/value - Despite the high proportion of remittances in international trade settlements, it has been hard to find studies that reflect the practical characteristics of remittances. This study classified the types of remittance into advance remittance, post remittance, and linked remittance, and built a trade payment prediction model by adding a letter of credit and mixed payment. In addition, the originality of this study is recognized in that a logistic model was constructed and meaningful results were derived.
https://doi.org/10.35611/jkt.2021.25.2.111 인용 PDF

시계열 분해 및 데이터 증강 기법 활용 건화물운임지수 예측 (Forecasting Baltic Dry Index by Implementing Time-Series Decomposition and Data Augmentation Techniques)

한민수;유성진
- 품질경영학회지
- /
- 제50권4호
- /
- pp.701-716
- /
- 2022
Purpose: This study aims to predict the dry cargo transportation market economy. The subject of this study is the BDI (Baltic Dry Index) time-series, an index representing the dry cargo transport market. Methods: In order to increase the accuracy of the BDI time-series, we have pre-processed the original time-series via time-series decomposition and data augmentation techniques and have used them for ANN learning. The ANN algorithms used are Multi-Layer Perceptron (MLP), Recurrent Neural Network (RNN), and Long Short-Term Memory (LSTM) to compare and analyze the case of learning and predicting by applying time-series decomposition and data augmentation techniques. The forecast period aims to make short-term predictions at the time of t+1. The period to be studied is from '22. 01. 07 to '22. 08. 26. Results: Only for the case of the MAPE (Mean Absolute Percentage Error) indicator, all ANN models used in the research has resulted in higher accuracy (1.422% on average) in multivariate prediction. Although it is not a remarkable improvement in prediction accuracy compared to uni-variate prediction results, it can be said that the improvement in ANN prediction performance has been achieved by utilizing time-series decomposition and data augmentation techniques that were significant and targeted throughout this study. Conclusion: Nevertheless, due to the nature of ANN, additional performance improvements can be expected according to the adjustment of the hyper-parameter. Therefore, it is necessary to try various applications of multiple learning algorithms and ANN optimization techniques. Such an approach would help solve problems with a small number of available data, such as the rapidly changing business environment or the current shipping market.
https://doi.org/10.7469/JKSQM.2022.50.4.701 인용 PDF KSCI

Building Knowledge Based Simulator for I11-Structured Dynamic Domain Using Cognitive Map : An Application to Stock Market Prediction

김현수
- 한국정보시스템학회지:정보시스템연구
- /
- 제2권
- /
- pp.127-140
- /
- 1993
PDF

국고채, 금리 스왑 그리고 통화 스왑 가격에 기반한 외환시장 환율예측 연구: 인공지능 활용의 실증적 증거 (A Study on Foreign Exchange Rate Prediction Based on KTB, IRS and CCS Rates: Empirical Evidence from the Use of Artificial Intelligence)

임현욱;정승환;이희수;오경주
- 지식경영연구
- /
- 제22권4호
- /
- pp.71-85
- /
- 2021
본 연구는 채권시장과 금리시장의 지표를 이용한 외환시장 환율예측 모델을 만드는데 있어 어떤 인공지능 방법론이 가장 적합한지 밝혀내는데 그 목적이 있다. 채권시장의 대표 상품인 국고채와 통안채는 위험회피 상황이 올 때 대규모로 매도되어지고 그런 경우 환율이 상승하는 모습을 자주 보여주었고, 금리시장에서 통화 스왑 (Cross Currency Swap) 가격은 달러 유동성 문제가 생길 때 주로 하락하였으며, 그 움직임은 환율의 상승에 직간접적인 영향을 미쳐온 점 등을 고려하면, 채권시장과 금리시장에서 거래되는 상품의 가격과 움직임은 외환시장에도 직간접적인 영향을 주고 있으며, 세 시장 사이엔 상호 유기적이고 보완적인 관계가 있다고 볼 수 있다. 지금까지 채권시장, 금리시장, 그리고 외환시장 사이의 관계와 연관성을 밝히는 연구는 있어왔으나, 과거 많은 환율예측 연구들이 주로 GDP, 경상수지 흑자/적자, 인플레이션 등 거시적인 지표를 기반으로 한 연구에 집중되어 왔으며, 채권시장과 금리시장 지표를 기반으로 인공지능을 활용하여 외환시장의 환율을 예측하는 적극적인 연구는 아직 진행되지 않았다. 본 연구는 채권시장 지표와 금리시장 지표를 기반으로, 비선형데이터 분석에 적합한 인공신경망(Artificial Neural Network) 모델과, 선형데이터 분석에 적합한 로지스틱 회귀분석 (Logistic regression), 그리고 비선형/선형데이터 분석에 활용 가능한 의사결정나무 (Decision Tree)를 각각 사용하여 환율예측 모델을 만들고 그 수익률을 비교하여 어떤 모델이 가장 외환시장 환율 예측을 하는데 적합한지 알려준다. 또한, 본 연구는 주식시장, 금리시장, 오일시장, 그리고 외환시장 환율 등 비선형적 시계열 데이터 분석에 많이 사용되어진 인공신경망 모델이 채권시장과 금리시장 지표를 기반으로 한 외환시장 환율예측 모델에 가장 적합한 방법론을 제공하고 있다는 것을 증명한다. 채권시장, 금리시장, 그리고 외환시장 간의 단순한 연관성을 밝히는 것을 넘어, 세 시장 간의 거래 신호를 포착하여 적극적인 상관관계를 밝히고 상호 유기적인 움직임을 증명하는 것은 단순히 외환시장 트레이더 들에게 새로운 트레이딩 모델을 제시하는 것뿐만 아니라 금융시장 전체의 효율성을 증가시키는데 기여할 것이라 기대한다.
https://doi.org/10.15813/kmr.2021.22.4.004 인용 PDF KSCI

지식 누적을 이용한 실시간 주식시장 예측 (A Real-Time Stock Market Prediction Using Knowledge Accumulation)

김진화;홍광헌;민진영
- 지능정보연구
- /
- 제17권4호
- /
- pp.109-130
- /
- 2011
연속발생 데이터는 데이터의 원천으로부터 데이터 저장소로 연속적으로 축적이 되는 데이터를 말한다. 이렇게 축적된 데이터의 크기는 시간이 지남에 따라 점점 커진다. 또한 이러한 대용량 데이터에서 정보를 추출하기 위해서는 저장공간, 시간, 그리고 많은 자원이 필요하다. 이러한 연속발생 데이터의 특성은 시간이 지남에 따라 축적된 대용량 데이터의 이용을 어렵고 고비용이 되게 한다. 만약 정보나 패턴을 추출할 때 누적된 전체 발생 데이터 중에서 최근의 일부만 사용 한다면 적은 일부 표본의 사용의 문제로 인하여 전체 데이터 사용에서 발견될 수 있는 유용한 정보의 유실이 있을 수 있다. 이러한 문제점을 해결하기 위해서 본 연구는 연속발생 데이터를 발생 시점에서 계속 모으기 보다 이러한 발생되는 데이터에서 규칙을 추출하여 효율적으로 지식을 관리하고자 한다. 이 방법은 기존의 방법에 비하여 적은 양의 데이터 저장공간을 필요로 한다. 또한 이렇게 축적된 규칙집합은 미래에 예측을 위해서 언제든 실시간 예측을 할 수 있게 준비가 된다. 여러 예측 모델을 결합시키는 방법인 앙상블 이론에 의하면 본 연구가 제시하는 데로 체계적으로 규칙집합을 시간에 따라 융합시킬 경우 더 나은 예측 성과가 가능하다. 본 연구는 주식시장의 변동성을 예측하기 위하여 주식시장 데이터를 사용하였다. 본 연구는 이 데이터를 이용해 본 연구가 제시하는 방법과 기존의 방법의 예측 정확도를 비교 하였다.
https://doi.org/10.13088/jiis.2011.17.4.109 인용 PDF KSCI

빅데이터를 활용한 인공지능 주식 예측 분석 (Stock prediction analysis through artificial intelligence using big data)

최훈
- 한국정보통신학회논문지
- /
- 제25권10호
- /
- pp.1435-1440
- /
- 2021
저금리 시대의 도래로 인해 많은 투자자들이 주식 시장으로 몰리고 있다. 과거의 주식 시장은 사람들이 기업 분석 및 각자의 투자기법을 통해 노동 집약적으로 주식 투자가 이루어졌다면 최근 들어 인공지능 및 데이터를 활용하여 주식 투자가 널리 이용되고 있는 실정이다. 인공지능을 통해 주식 예측의 성공률은 현재 높지 않아 다양한 인공지능 모델을 통해 주식 예측률을 높이는 시도를 하고 있다. 본 연구에서는 다양한 인공지능 모델에 대해 살펴보고 각 모델들간의 장단점 및 예측률을 파악하고자 한다. 이를 위해, 본 연구에서는 주식예측 인공지능 프로그램으로 인공신경망(ANN), 심층 학습 또는 딥 러닝(DNN), k-최근접 이웃 알고리즘(k-NN), 합성곱 신경망(CNN), 순환 신경망(RNN), LSTM에 대해 살펴보고자 한다.
https://doi.org/10.6109/jkiice.2021.25.10.1435 인용 PDF KSCI

통계적 및 인공지능 모형 기반 태양광 발전량 예측모델 비교 및 재생에너지 발전량 예측제도 정산금 분석 (Comparison of solar power prediction model based on statistical and artificial intelligence model and analysis of revenue for forecasting policy)

이정인;박완기;이일우;김상하
- 전기전자학회논문지
- /
- 제26권3호
- /
- pp.355-363
- /
- 2022
우리나라는 2050년 탄소중립을 목표로 신재생에너지 중심으로 에너지 공급원을 전환하고 확대하는 계획을 추진 중이다. 신재생에너지의 간헐적 특성으로 에너지 공급이 불안정성이 커짐에 따라 정확한 신재생에너지 발전량 예측의 중요성이 함께 커지고 있다. 이에 따라 정부는 신재생에너지를 집합화하여 관리하기 위한 소규모 전력중개시장을 개설하였고, 재생에너지 발전량 예측제도를 도입하여 예측정확도에 따라 정산금을 지급하는 제도를 시행 중이다. 본 논문에서는 우리나라 신재생에너지 전원의 대부분을 차지하는 태양광 발전에 대하여 통계적 및 인공지능 모형을 이용하여 예측모델을 구현하였으며, 각 모형의 예측정확도 결과를 비교 분석하였다. 비교 모델 중에서 CNN-LSTM(Convolutional Long Short-Term Memory Neural Networks) 모형이 가장 높은 성능을 가짐을 확인하였다. 예측정확도에 따른 예측제도 정산금 수익을 추정해보았고, 예측보유 기술 수준에 따라 수익 편차가 24% 정도 커질 수 있음을 확인하였다.
https://doi.org/10.7471/ikeee.2022.26.3.355 인용 PDF KSCI

그래디언트 부스팅을 활용한 암호화폐 가격동향 예측 (Prediction of Cryptocurrency Price Trend Using Gradient Boosting)

허주성;권도형;김주봉;한연희;안채헌
- 정보처리학회논문지:소프트웨어 및 데이터공학
- /
- 제7권10호
- /
- pp.387-396
- /
- 2018
과거부터 주식시장의 주가 예측은 풀리지 않는 난제이다. 이를 과학적으로 예측하기 위해 다양한 시도 및 연구들이 있어왔지만 정확한 가격을 예측하는 것은 불가능하다. 최근 분산 원장이라는 개념을 기술적으로 구현한 최초의 암호화폐인 비트코인을 시작으로 다양한 종류의 암호화폐가 개발되면서 암호화폐 시장이 형성되었고, 그 가격을 예측하기 위해 다양한 접근들이 시도되고 있다. 특히, 기존의 전통적인 주식시장에서의 주가 예측 기법들을 적용하려는 시도부터 딥러닝과 강화학습을 적용하려는 시도까지 다양하다. 하지만 암호화폐 시장은 기존 주식 시장에는 없던 여러 가지 새로운 특징을 가지는 시장으로서 전통적인 주식 시장 분석 기술뿐만 아니라 암호화폐 시장에 적합한 새로운 분석 기술에 관한 수요가 증가하고 있는 상황이다. 본 연구에서는 우선 빗썸의 API를 통하여 7개의 암호화폐 가격 데이터를 수집 및 가공하였다. 이후, Data-Driven 방식의 지도학습 기반 기계학습 모델인 그래디언트 부스팅 모델을 채택하여 암호화폐 가격 데이터 변화를 학습하고, 검증단계에서 가장 최적의 모델 파라미터를 산출하고, 최종적으로 테스트 데이터를 활용하여 암호화폐 가격동향 예측 성능을 평가한다.
https://doi.org/10.3745/KTSDE.2018.7.10.387 인용 PDF KSCI

검색결과 528건 처리시간 0.027초

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

자세히 찾기

이미지 검색 (β)