Similarity/dissimilarity calculation methods of DNA sequences: A survey

Xin Jin, Qian Jiang, Yanyan Chen, Sj Lee, Rencan Nie, Shaowen Yao*, Dongming Zhou, Kangjian He

*Corresponding author for this work

Research output: Contribution to journalShort surveypeer-review

22 Scopus citations


DNA sequence similarity/dissimilarity analysis is a fundamental task in computational biology, which is used to analyze the similarity of different DNA sequences for learning their evolutionary relationships. In past decades, a large number of similarity analysis methods for DNA sequence have been proposed due to the ever-growing demands. In order to learn the advances of DNA sequence similarity analysis, we make a survey and try to promote the development of this field. In this paper, we first introduce the related knowledge of DNA similarities analysis, including the data sets, similarities distance and output data. Then, we review recent algorithmic developments for DNA similarity analysis to represent a survey of the art in this field. At last, we summarize the corresponding tendencies and challenges in this research field. This survey concludes that although various DNA similarity analysis methods have been proposed, there still exist several further improvements or potential research directions in this field.

Original languageEnglish
Pages (from-to)342-355
Number of pages14
JournalJournal of Molecular Graphics and Modelling
StatePublished - Sep 2017


  • DNA sequence analysis
  • Evolutionary relationship
  • Feature extraction
  • Graphical representation
  • Similarity analysis


Dive into the research topics of 'Similarity/dissimilarity calculation methods of DNA sequences: A survey'. Together they form a unique fingerprint.

Cite this