Relevance Feature Selection with Data Cleaning for Intrusion Detection System

被引:0
|
作者
Suthaharan, Shan [1 ]
Panchagnula, Tejaswi [1 ]
机构
[1] Univ N Carolina, Dept Comp Sci, Greensboro, NC 27412 USA
关键词
intrusion detection; Rough Set Theory; labeled datasets; NSL-KDD dataset; relevance feature selection;
D O I
暂无
中图分类号
TM [电工技术]; TN [电子技术、通信技术];
学科分类号
0808 ; 0809 ;
摘要
Labeled datasets play a major role in the process of validating and evaluating machine learning techniques in intrusion detection systems. In order to obtain good accuracy in the evaluation, very large datasets should be considered. Intrusion traffic and normal traffic are in general dependent on a large number of network characteristics called features. However not all of these features contribute to the traffic characteristics. Therefore, eliminating the non-contributing features from the datasets, to facilitate speed and accuracy to the evaluation of machine learning techniques, becomes an important requirement. In this paper we suggest an approach which analyzes the intrusion datasets, evaluates the features for its relevance to a specific attack, determines the level of contribution of feature, and eliminates it from the dataset automatically. We adopt the Rough Set Theory (RST) based approach and select relevance features using multidimensional scatter-plot automatically. A pair-wise feature selection process is adopted to simplify. In our previous research we used KDD'99 dataset and validated the RST based approach. There are lots of redundant data entries in KDD'99 and thus the machine learning techniques are biased towards most occurring events. This property leads the algorithms to ignore less frequent events which can be more harmful than most occurring events. False positives are another important drawback in KDD'99 dataset. In this paper, we adopt NSL-KDD dataset (an improved version of KDD'99 dataset) and validate the automated RST based approach. The approach presented in this paper leads to a selection of most relevance features and we expect that the intrusion detection research using KDD'99-based datasets will benefit from the good understanding of network features and their influences to attacks.
引用
收藏
页数:6
相关论文
共 50 条
  • [31] Mutual clustered redundancy assisted feature selection for an intrusion detection system
    Veeranna, T.
    Reddi, Kiran Kumar
    [J]. JOURNAL OF HIGH SPEED NETWORKS, 2022, 28 (04) : 257 - 273
  • [32] An IWD-based feature selection method for intrusion detection system
    Neha Acharya
    Shailendra Singh
    [J]. Soft Computing, 2018, 22 : 4407 - 4416
  • [33] INTRUSION DETECTION SYSTEM BASED ON FEATURE SELECTION AND SUPPORT VECTOR MACHINE
    Zhang Xue-qin
    Gu Chun-hua
    Lin Jia-jun
    [J]. 2006 FIRST INTERNATIONAL CONFERENCE ON COMMUNICATIONS AND NETWORKING IN CHINA, 2006,
  • [34] Feature Subset Selection Using Genetic Algorithm for Intrusion Detection System
    Behjat, Amir Rajabi
    Vatankhah, Najmeh
    Mustapha, Aida
    [J]. ADVANCED SCIENCE LETTERS, 2014, 20 (01) : 235 - 238
  • [35] Evolutionary Algorithm-based Feature Selection for an Intrusion Detection System
    Singh, Devendra Kumar
    Shrivastava, Manish
    [J]. ENGINEERING TECHNOLOGY & APPLIED SCIENCE RESEARCH, 2021, 11 (03) : 7130 - 7134
  • [36] Intrusion Detection System using Bayesian Network and Feature Subset Selection
    Jabbar, M. A.
    Aluvalu, Rajanikanth
    Reddy, S. Sai Satyanarayana
    [J]. 2017 IEEE INTERNATIONAL CONFERENCE ON COMPUTATIONAL INTELLIGENCE AND COMPUTING RESEARCH (ICCIC), 2017, : 640 - 644
  • [37] Utilizing Feature Selection Techniques in Intrusion Detection System for Internet of Things
    Jafier, Shatha H.
    [J]. ICFNDS'18: PROCEEDINGS OF THE 2ND INTERNATIONAL CONFERENCE ON FUTURE NETWORKS AND DISTRIBUTED SYSTEMS, 2018,
  • [38] An IWD-based feature selection method for intrusion detection system
    Acharya, Neha
    Singh, Shailendra
    [J]. SOFT COMPUTING, 2018, 22 (13) : 4407 - 4416
  • [39] Majority Voting and Feature Selection Based Network Intrusion Detection System
    Patil, Dharmaraj R.
    Pattewar, Tareek M.
    [J]. EAI ENDORSED TRANSACTIONS ON SCALABLE INFORMATION SYSTEMS, 2022, 9 (06):
  • [40] A FEATURE SELECTION ALGORITHM DESIGN AND ITS IMPLEMENTATION IN INTRUSION DETECTION SYSTEM
    杨向荣
    沈钧毅
    [J]. Journal of Pharmaceutical Analysis, 2003, (02) : 134 - 138