Inside the Pan-genome - Methods and Software Overview

Luis      Carlos Guimaraes; Leandro      Benevides de Jesus; Marcus      Vinicius Canario Viana; Artur      Silva; Rommel      Thiago Juca Ramos; Siomar      de Castro Soares; Vasco      Azevedo

doi:10.2174/1389202916666150423002311

Abstract

The number of genomes that have been deposited in databases has increased exponentially after the advent of Next-Generation Sequencing (NGS), which produces high-throughput sequence data; this circumstance has demanded the development of new bioinformatics software and the creation of new areas, such as comparative genomics. In comparative genomics, the genetic content of an organism is compared against other organisms, which helps in the prediction of gene function and coding region sequences, identification of evolutionary events and determination of phylogenetic relationships. However, expanding comparative genomics to a large number of related bacteria, we can infer their lifestyles, gene repertoires and minimal genome size. In this context, a powerful approach called Pan-genome has been initiated and developed. This approach involves the genomic comparison of different strains of the same species, or even genus. Its main goal is to establish the total number of non-redundant genes that are present in a determined dataset. Pan-genome consists of three parts: core genome; accessory or dispensable genome; and species-specific or strain-specific genes. Furthermore, pan-genome is considered to be “open” as long as new genes are added significantly to the total repertoire for each new additional genome and “closed” when the newly added genomes cannot be inferred to significantly increase the total repertoire of the genes. To perform all of the required calculations, a substantial amount of software has been developed, based on orthologous and paralogous gene identification.

Keywords: Pan-genome, Core genome, Accessory genome, Species-specific genome, Comparative genome.

« Previous Next »

Graphical Abstract

Rights & Permissions Print Cite

Article Metrics

31

1

Journal Information

For Authors

For Editors

For Reviewers

Explore Articles

Open Access

Open Access Articles

For Visitors

DOI https://dx.doi.org/10.2174/1389202916666150423002311	Print ISSN 1389-2029
Publisher Name Bentham Science Publisher	Online ISSN 1875-5488

Current Genomics

Inside the Pan-genome - Methods and Software Overview

Abstract

Graphical Abstract

Advanced AI Techniques in Big Genomic Data Analysis

Current Genomics in Cardiovascular Research

Genomic Insights into Oncology: Harnessing Machine Learning for Breakthroughs in Cancer Genomics.

Integrating Artificial Intelligence and Omics Approaches in Complex Diseases

Current Genomics

Inside the Pan-genome - Methods and Software Overview

Abstract

Graphical Abstract

Call for Papers in Thematic Issues

Advanced AI Techniques in Big Genomic Data Analysis

Current Genomics in Cardiovascular Research

Genomic Insights into Oncology: Harnessing Machine Learning for Breakthroughs in Cancer Genomics.

Integrating Artificial Intelligence and Omics Approaches in Complex Diseases

Related Journals

Related Books

Related Articles