통합 검색 | Korea Science

The Adaptive SPAM Mail Detection System using Clustering based on Text Mining

Hong, Sung-Sam;Kong, Jong-Hwan;Han, Myung-Mook
- KSII Transactions on Internet and Information Systems (TIIS)
- /
- 제8권6호
- /
- pp.2186-2196
- /
- 2014
Spam mail is one of the most general mail dysfunctions, which may cause psychological damage to internet users. As internet usage increases, the amount of spam mail has also gradually increased. Indiscriminate sending, in particular, occurs when spam mail is sent using smart phones or tablets connected to wireless networks. Spam mail consists of approximately 68% of mail traffic; however, it is believed that the true percentage of spam mail is at a much more severe level. In order to analyze and detect spam mail, we introduce a technique based on spam mail characteristics and text mining; in particular, spam mail is detected by extracting the linguistic analysis and language processing. Existing spam mail is analyzed, and hidden spam signatures are extracted using text clustering. Our proposed method utilizes a text mining system to improve the detection and error detection rates for existing spam mail and to respond to new spam mail types.
https://doi.org/10.3837/tiis.2014.06.022 인용 PDF KSCI KPUBS HTML

Improved Spam Filter via Handling of Text Embedded Image E-mail

Youn, Seongwook;Cho, Hyun-Chong
- Journal of Electrical Engineering and Technology
- /
- 제10권1호
- /
- pp.401-407
- /
- 2015
The increase of image spam, a kind of spam in which the text message is embedded into attached image to defeat spam filtering technique, is a major problem of the current e-mail system. For nearly a decade, content based filtering using text classification or machine learning has been a major trend of anti-spam filtering system. Recently, spammers try to defeat anti-spam filter by many techniques. Text embedding into attached image is one of them. We proposed an ontology spam filters. However, the proposed system handles only text e-mail and the percentage of attached images is increasing sharply. The contribution of the paper is that we add image e-mail handling capability into the anti-spam filtering system keeping the advantages of the previous text based spam e-mail filtering system. Also, the proposed system gives a low false negative value, which means that user's valuable e-mail is rarely regarded as a spam e-mail.
https://doi.org/10.5370/JEET.2015.10.1.401 인용 PDF KSCI KPUBS HTML

URL 빈도분석을 이용한 스팸메일 차단 방법 (A spam mail blocking method using URL frequency analysis)

백기영;이철수;류재철
- 정보보호학회논문지
- /
- 제14권6호
- /
- pp.135-148
- /
- 2004
최근 다양하게 변하는 스팸메일은 단어에 의한 기존의 스팸메일 판별 방법으로는 차단하기 어렵다. 이와 같은 문제를 해결하고자 URL 빈도분석을 이용한 스팸메일 관별 규칙 생성 방법을 제안한다. 제안한 방법은 스팸메일을 수집하고, 수집된 스팸메일에서 특징이 되는 URL을 추출하고, 이를 정규화하여 시간 빈도에 따른 스팸메일 판별 규칙 생성하여 스팸메일을 차단하는 단계로 구성된다. 이는 다양한 스팸메일에 대응할 수 있으며 변화하는 스팸메일의 형태에 대해서도 대응할 수 있는 구조를 가지고 있다.
https://doi.org/10.13089/JKIISC.2004.14.6.135 인용 PDF KSCI HTML

An Architecture for Certificate and Agent Based E-mailing to Block Spam Mail

Nam, Sang-Zo
- 지능정보연구
- /
- 제9권2호
- /
- pp.39-50
- /
- 2003
Deleting unsolicited email, popularly known as spam mail, is an annoying task for Internet users. Moreover, spam mail causes a variety of social problems. At present, legal restrictions cannot eradicate spam senders. As a result, many technical methods to eliminate spam mail such as spam filtering and online stamps have been introduced. However, the process of blocking spam mail can inadvertently result in suspension of indispensable or beneficial communication. In this paper, we propose a certificate and agent based emailing architecture that can block spam mail, while at the same time approve certified mail. This architecture can be accelerated by synergistic utilization of digital signature and electronic document interchange.
PDF

A Proposed Architecture for Certificate and Agent Based E-mailing to Block Spam Mail

Nam, Sang-Zo
- 한국산학기술학회:학술대회논문집
- /
- 한국산학기술학회 2003년도 Proceeding
- /
- pp.28-34
- /
- 2003
Deleting unsolicited email, popularly known as spam mail, is an annoying task for Internet users. Moreover, spam mail causes a variety of social problems. At present, legal restrictions cannot eradicate spam senders. As a result many technical methods to eliminate spam mail such as spam filtering and online stamps have been introduced. However, the process of blocking spam mail can inadvertently result in suspension of indispensable or beneficial communication. In this paper, we propose a certificate and agent based emailing architecture that can block spam mail, while at the same time approve certified mail. This architecture can be accelerated by synergistic utilization of digital signature and electronic document interchange.
PDF

Analyzing the correlation of Spam Recall and Thesaurus

Kang, Sin-Jae;Kim, Jong-Wan
- 한국정보기술응용학회:학술대회논문집
- /
- 한국정보기술응용학회 2005년도 6th 2005 International Conference on Computers, Communications and System
- /
- pp.21-25
- /
- 2005
In this paper, we constructed a two-phase spam-mail filtering system based on the lexical and conceptual information. There are two kinds of information that can distinguish the spam mail from the legitimate mail. The definite information is the mail sender's information, URL, a certain spam list, and the less definite information is the word list and concept codes extracted from the mail body. We first classified the spam mail by using the definite information, and then used the less definite information. We used the lexical information and concept codes contained in the email body for SVM learning in the $2^{nd}$ phase. According to our results the spam precision was increased if more lexical information was used as features, and the spam recall was increased when the concept codes were included in features as well.
PDF

Analyzing the Effect of Lexical and Conceptual Information in Spam-mail Filtering System

Kang Sin-Jae;Kim Jong-Wan
- International Journal of Fuzzy Logic and Intelligent Systems
- /
- 제6권2호
- /
- pp.105-109
- /
- 2006
In this paper, we constructed a two-phase spam-mail filtering system based on the lexical and conceptual information. There are two kinds of information that can distinguish the spam mail from the ham (non-spam) mail. The definite information is the mail sender's information, URL, a certain spam keyword list, and the less definite information is the word list and concept codes extracted from the mail body. We first classified the spam mail by using the definite information, and then used the less definite information. We used the lexical information and concept codes contained in the email body for SVM learning in the 2nd phase. According to our results the ham misclassification rate was reduced if more lexical information was used as features, and the spam misclassification rate was reduced when the concept codes were included in features as well.
https://doi.org/10.5391/IJFIS.2006.6.2.105 인용 PDF KSCI

The Exploratory Analysis for Spam Mail Data Using Correspondence Analysis

Shin, Yang-Kyu
- Journal of the Korean Data and Information Science Society
- /
- 제16권4호
- /
- pp.735-744
- /
- 2005
The number of electronic mail(E-mail) has been increased dramatically as a result of expanding internet and information technology. Although there are many conveniences of E-mail in the bright side, some serious problems occur because of E-mail in its dark side. One of the problems is spam-mail which is unsolicited mail and also called bulk mail. This paper presents a set of patterns of spam-mail occurrences within a week using the correspondence analysis. The correspondence analysis is an exploratory multivariate technique that converts data into a particular type of graphical display in which the rows and columns are depicted as points. One of the meaningful patterns is a great increment of adult and phishing related spam-mails at weekends so any spam-mail filters should be designed to cope with this pattern.
PDF

수집과 빈도분석을 통한 스팸메일 차단 방법 (A spam mail blocking method using collection and frequency analysis)

백기영;김승해;최장원;류재철
- 정보처리학회논문지C
- /
- 제12C권1호
- /
- pp.137-146
- /
- 2005
인터넷을 이용한 이메일은 이제 소수의 통신 수단이 아닌 일반인이 널리 사용하는 기본적인 통신 수단으로 자리잡고 있으며, 이에 따른 스팸메일의 피해 규모도 날로 커지고 있다. 현재 다양한 방법의 스팸메일 차단 방법이 제안되고 수행되고 있으나 다양해지는 스팸메일에 대응하기에는 역부족이다. 이 논문에서 제시한 스팸메일 차단 방은은 수집, 빈도분석과 차단의 3단계로 구성되며, 수집되는 스팸메일을 이용하여 다양한 스팸메일에 대응할 수 있으며, 변화하는 스팸메일의 형태에 대해서도 대응할 수 있는 구조를 가지고 있다.
https://doi.org/10.3745/KIPSTC.2005.12C.1.137 인용 PDF KSCI

스팸메일 방지를 위한 제도적 기술적 해결방안에 관한 연구 (The Study about Solution for The Protection of Spam Mails)

강장묵;유의상;이정훈
- 한국IT서비스학회지
- /
- 제2권1호
- /
- pp.25-34
- /
- 2003
Spam mail is one of the side effect of the development and improvement of the internet that restrains the privacy of the individual on line. However indiscriminate application of Spam mail blocking can also cause significant violation on freedom of doing business to the fluent commercial transactions on line. Therefore this research looks at the exact understanding of the concept of Spam mail and inquiry on Its issues. Also it looks at the case studies of its institutional solutions in USA and Europe as well as the advantage and disadvantage of the case studies on its technical solution. Finally, the research inquires into overall prevention of Spam mail, which considers both technical and institutional soiution. With this research, limitations of current Spam mail prevention system and technology are pointed out and more effective course of overall Spam mail prevention solution is studied.
PDF KSCI

검색결과 114건 처리시간 0.022초

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

자세히 찾기

이미지 검색 (β)