Using simulation to evaluate prediction techniques

被引:0
|
作者
Shepperd, M [1 ]
Kadoda, G [1 ]
机构
[1] Bournemouth Univ, Sch Design Engn & Comp, Empirical Software Engn Res Grp, Poole BH12 5BB, Dorset, England
关键词
prediction system; simulation; dataset characteristic;
D O I
暂无
中图分类号
TP31 [计算机软件];
学科分类号
081202 ; 0835 ;
摘要
The need for accurate software prediction systems increases as software becomes much larger and more complex. A variety of techniques have been proposed, however, none has proved consistently accurate and there is still much uncertainty as to what technique suits which type of prediction problem. We believe that the underlying characteristics - size, number of features, type of distribution, etc. - of the dataset influence the choice of the prediction system to be used. In previous work, it has proved difficult to obtain significant results over small datasets. Consequently we required large validation datasets, moreover, we wished to control the characteristics of such datasets in order to systematically explore the relationship between accuracy, choice of prediction system and dataset characteristic. Our solution has been to simulate data allowing both control and the possibility of large (1000) validation cases. In this paper we compared regression, rule induction and nearest neighbour (a form of case based reasoning). The results suggest that there are significant differences depending upon the characteristics of the dataset. Consequently researchers should consider prediction context when evaluating competing prediction systems. We also observed that the more "messy" the data and the more complex the relationship with the dependent variable the more variability in the results. This became apparent since we sampled two different training sets from each simulated population of data. In the more complex cases we observed significantly different results depending upon the training set. This suggests that researchers will need to exercise caution when comparing different approaches and utilise procedures such as bootstrapping in order to generate multiple samples for training purposes.
引用
收藏
页码:349 / 359
页数:11
相关论文
共 50 条
  • [1] Using Simulation Models to Evaluate Ape Nest Survey Techniques
    Boyko, Ryan H.
    Marshall, Andrew J.
    [J]. PLOS ONE, 2010, 5 (05):
  • [2] Comparing software prediction techniques using simulation
    Shepperd, M
    Kadoda, G
    [J]. IEEE TRANSACTIONS ON SOFTWARE ENGINEERING, 2001, 27 (11) : 1014 - 1022
  • [3] Prediction of Casting Defects Using Simulation Techniques
    Rodriguez Moliner, Tania
    Parada Exposito, Andres
    Ordonez Hernandez, Urbano
    [J]. REVISTA CUBANA DE INGENIERIA, 2010, 1 (02): : 65 - 69
  • [4] Taking Another Look: Using New Simulation Techniques to Evaluate Old Antibiotics
    Cattrall, J. W. S.
    Asin-Prieto, E.
    Freeman, J.
    Troconiz, I. F.
    Kirby, A.
    [J]. JOURNAL OF PATHOLOGY, 2019, 249 : S22 - S22
  • [5] CALCULATING THE PRECISION OF PREDICTION OF HAUTEUR OF A TOP USING SIMULATION TECHNIQUES
    DAVID, HG
    [J]. WOOL TECHNOLOGY AND SHEEP BREEDING, 1993, 41 (01): : 56 - 61
  • [6] USING SIMULATION TO EVALUATE AS/RS
    ERICKSON, CJ
    [J]. MANUFACTURING ENGINEERING, 1987, 99 (03): : 55 - 55
  • [7] Development of a simulation result management and prediction system using machine learning techniques
    Lee, Ki Yong
    Suh, Young-Kyoon
    Cho, Kum Won
    [J]. INTERNATIONAL JOURNAL OF DATA MINING AND BIOINFORMATICS, 2017, 19 (01) : 75 - 96
  • [8] Using Simulation to Evaluate Nurse Competencies
    Vanderzwan, Kathryn J.
    Schwind, Julie
    Obrecht, Jennifer
    O'Rourke, Jennifer
    Johnson, Alexia Hieber
    [J]. JOURNAL FOR NURSES IN PROFESSIONAL DEVELOPMENT, 2020, 36 (03) : 163 - 166
  • [9] MOLECULAR SIMULATION TECHNIQUES AND STRUCTURE PREDICTION.
    Freeman, Clive M.
    Gorman, Alan M.
    Levine, Steve M.
    Leusen, Frank J. J.
    Deem, Michael W.
    Falcioni, Marco
    Newsam, John M.
    [J]. ACTA CRYSTALLOGRAPHICA A-FOUNDATION AND ADVANCES, 1999, 55 : 214 - 215
  • [10] Geology Prediction Techniques for Reservoir Evolution Simulation
    Q. Wendao
    Y. Taiju
    Zh. Changmin
    H. Guowei
    H. Miao
    X. Min
    Y. Xiujin
    Y. Lan
    [J]. Geotectonics, 2019, 53 : 399 - 418