A novel graphical representation and similarity analysis of protein sequences based on physicochemical properties

被引:13
|
作者
Mahmoodi-Reihani, Mehri [1 ]
Abbasitabar, Fatemeh [2 ]
Zare-Shahabadi, Vahid [1 ]
机构
[1] Islamic Azad Univ, Mahshahr Branch, Dept Chem, Mahshahr 6351977439, Iran
[2] Islamic Azad Univ, Marvdasht Branch, Dept Chem, Marvdasht 7371113119, Iran
关键词
Protein sequence; Graphical representation; Principal component analysis; Physicochemical property; Moving window correlation coefficient; ACID INDEX DATABASE; AMINO-ACID; DNA-SEQUENCES; NUMERICAL CHARACTERIZATION; 2D; ALIGNMENT; MATRIX; SIMILARITY/DISSIMILARITY; AAINDEX;
D O I
10.1016/j.physa.2018.07.011
中图分类号
O4 [物理学];
学科分类号
0702 ;
摘要
One of popular topic in bioinformatics is protein sequence analysis. The graphical representation of protein sequence is a simple and common way to visualize protein sequences. In this study, a numerical descriptive vector for a given protein sequence is calculated based on twelve physicochemical properties of amino acids (AAs) and principal component analysis (PCA). Each entry of the descriptive vector corresponds to one AA in the sequence. By this vector, an intuitive spectrum-like graphical representation of protein sequence is proposed. Squared correlation coefficient as well as moving window correlation coefficient, as a new similarity/dissimilarity measure, were used to compare different sequences. Applicability of the proposed method is assessed by analyzing the nine ND5 proteins. The results revealed the utility of the proposed method. (C) 2018 Elsevier B.V. All rights reserved.
引用
收藏
页码:477 / 485
页数:9
相关论文
共 50 条
  • [31] A new graphical representation and its application in similarity/dissimilarity analysis of DNA sequences
    Luo, Jiawei
    Guo, Jiachen
    Li, Yang
    2010 4TH INTERNATIONAL CONFERENCE ON BIOINFORMATICS AND BIOMEDICAL ENGINEERING (ICBBE 2010), 2010,
  • [32] On the similarity/dissimilarity of DNA sequences based on 4D graphical representation
    TANG XiaoChanZHOU PanPan QIU WenYuan Department of ChemistryState Key Laboratory of Applied Organic ChemistryLanzhou UniversityLanzhou China
    Chinese Science Bulletin, 2010, 55 (08) : 701 - 704
  • [33] On the similarity/dissimilarity of DNA sequences based on 4D graphical representation
    TANG XiaoChan
    Science Bulletin, 2010, (08) : 701 - 704
  • [34] On the similarity/dissimilarity of DNA sequences based on 4D graphical representation
    Tang XiaoChan
    Zhou PanPan
    Qiu WenYuan
    CHINESE SCIENCE BULLETIN, 2010, 55 (08): : 701 - 704
  • [35] A new 3D graphical representation for similarity/dissimilarity studies of protein sequences
    Chen, Yan
    Li, Kang-Shun
    Chang, Shan
    Yang, Lei
    Computer Modelling and New Technologies, 2014, 18 (12): : 296 - 303
  • [36] Novel techniques of graphical representation and analysis of DNA sequences - A review
    Roy, A
    Raychaudhury, C
    Nandy, A
    JOURNAL OF BIOSCIENCES, 1998, 23 (01) : 55 - 71
  • [37] Novel techniques of graphical representation and analysis of DNA sequences—A review
    A. Roy
    C. Raychaudhury
    A. Nandy
    Journal of Biosciences, 1998, 23 : 55 - 71
  • [38] A Novel Method for Similarity Analysis of Protein Sequences
    Liu, Longlong
    Zhao, Tingting
    Liu, Maojuan
    PROCEEDINGS OF THE 5TH INTERNATIONAL CONFERENCE ON ADVANCED DESIGN AND MANUFACTURING ENGINEERING, 2015, 39 : 2216 - 2220
  • [39] Hydropathy and Conformational Similarity-Based Distributed Representation of Protein Sequences for Properties Prediction
    Hrushikesh Bhosale
    Ashwin Lahorkar
    Divye Singh
    Aamod Sane
    Jayaraman Valadi
    SN Computer Science, 2022, 3 (1)
  • [40] Novel graphical representation of genome sequence and its applications in similarity analysis
    Yu, Hong-Jie
    Huang, De-Shuang
    PHYSICA A-STATISTICAL MECHANICS AND ITS APPLICATIONS, 2012, 391 (23) : 6128 - 6136