Analysis of multilocus fingerprinting data sets containing missing data

被引:357
|
作者
Schlueter, Philipp M.
Harris, Stephen A.
机构
[1] Univ Vienna, Inst Bot, Dept Systemat & Evolutionary Bot, A-1030 Vienna, Austria
[2] Univ Oxford, Dept Plant Sci, Oxford OX1 3RB, England
来源
MOLECULAR ECOLOGY NOTES | 2006年 / 6卷 / 02期
关键词
DNA fingerprinting; dominant markers; Jaccard's similarity coefficient; missing data; Shannon's diversity index;
D O I
10.1111/j.1471-8286.2006.01225.x
中图分类号
Q5 [生物化学]; Q7 [分子生物学];
学科分类号
071010 ; 081704 ;
摘要
Missing data are commonly encountered using multilocus, fragment-based (dominant) fingerprinting methods, such as random amplified polymorphic DNA (RAPD) or amplified fragment length polymorphism (AFLP). Data sets containing missing data have been analysed by eliminating those bands or samples with missing data, assigning values to missing data or ignoring the problem. Here, we present a method that uses random assignments of band presence-absence to the missing data, implemented by the computer program FAMD (available from http://homepage.univie.ac.at/philipp.maria.schlueter/famd.html), for analyses based on pairwise similarity and Shannon's index. When missing values group in a data set, sample or band elimination is likely to be the most appropriate action. However, when missing values are scattered across the data set, minimum, maximum and average similarity coefficients are a simple means of visualizing the effects of missing data on tree structure. Our approach indicates the range of values that a data set containing missing data points might generate, and forces the investigator to consider the effects of missing values on data interpretation.
引用
收藏
页码:569 / 572
页数:4
相关论文
共 50 条
  • [31] Missing data: Implications for analysis
    Fitzmaurice, Garrett
    NUTRITION, 2008, 24 (02) : 200 - 202
  • [32] INTERVENTION ANALYSIS WITH MISSING DATA
    LETTENMAIER, DP
    TRANSACTIONS-AMERICAN GEOPHYSICAL UNION, 1978, 59 (04): : 282 - 282
  • [33] SOLAS For Missing Data Analysis
    Bernaards, Coen A.
    STRUCTURAL EQUATION MODELING-A MULTIDISCIPLINARY JOURNAL, 1999, 6 (03) : 301 - 304
  • [34] FACTOR ANALYSIS WITH MISSING DATA
    WOODBURY, MA
    SILER, W
    ANNALS OF THE NEW YORK ACADEMY OF SCIENCES, 1966, 128 (A3) : 746 - &
  • [35] Data Reduction Analysis for Climate Data Sets
    Liu, Songbin
    Huang, Xiaomeng
    Fu, Haohuan
    Yang, Guangwen
    Song, Zhenya
    INTERNATIONAL JOURNAL OF PARALLEL PROGRAMMING, 2015, 43 (03) : 508 - 527
  • [36] Data Reduction Analysis for Climate Data Sets
    Songbin Liu
    Xiaomeng Huang
    Haohuan Fu
    Guangwen Yang
    Zhenya Song
    International Journal of Parallel Programming, 2015, 43 : 508 - 527
  • [37] Missing data mechanisms and their implications on the analysis of categorical data
    Frederico Z. Poleto
    Julio M. Singer
    Carlos Daniel Paulino
    Statistics and Computing, 2011, 21 : 31 - 43
  • [38] Missing phenotype data imputation in pedigree data analysis
    Fridley, B
    de Andrade, M
    GENETIC EPIDEMIOLOGY, 2005, 29 (03) : 249 - 249
  • [39] Missing data mechanisms and their implications on the analysis of categorical data
    Poleto, Frederico Z.
    Singer, Julio M.
    Paulino, Carlos Daniel
    STATISTICS AND COMPUTING, 2011, 21 (01) : 31 - 43
  • [40] Analysis of repeated binary data: sensitivity to missing data
    Minini, P
    Chavance, M
    REVUE D EPIDEMIOLOGIE ET DE SANTE PUBLIQUE, 2004, 52 (05): : 455 - 464