Eukaryotes evade information storage-replication rate trade-off with endosymbiont assistance leading to larger genomes
Genome length varies widely among organisms, from compact genomes of prokaryotes to vast and complex genomes of eukaryotes. In this study, we theoretically identify the evolutionary pressures that may have driven this divergence in genome length. We use a parameter-free model to study genome length evolution under selection pressure to minimize replication time and maximize information storage capacity. We show that prokaryotes tend to reduce genome length, constrained by a single replication origin, while eukaryotes expand their genomes by incorporating multiple replication origins. We propose a connection between genome length and cellular energetics, suggesting that endosymbiotic organelles, mitochondria and chloroplasts, evolutionarily regulate the number of replication origins, thereby influencing genome length in eukaryotes. We show that the above two selection pressures also lead to strict equalization of the number of purines and their corresponding base-pairing pyrimidines within a single DNA strand, known as Chagraff's second parity rule, a hitherto unexplained observation in genomes of nearly all known species. This arises from the symmetrization of replichore length, another observation that has been shown to hold across species, which our model reproduces. The model also reproduces other experimentally observed phenomena, such as a general preference for deletions over insertions, and elongation and high variance of genome lengths under reduced selection pressure for replication rate, termed the C-value paradox. We highlight the possibility of regulation of the firing of latent replication origins in response to cues from the extracellular environment leading to the regulation of cell cycle rates in multicellular eukaryotes.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Genome replication in asynchronously growing microbial populations
Biological cells replicate their genomes in a well-planned manner. The DNA replication program of an organism determines the timing at which different genomic regions are replicated, with fundamental consequences for cel…
Computational prediction of replication sites in DNA sequences using complex number representation
Computational prediction of origin of replication (ORI) has been of great interest in bioinformatics and several methods including GC-skew, auto-correlation etc. have been explored in the past. In this paper, we have ext…
iCorr : Complex correlation method to detect origin of replication in prokaryotic and eukaryotic genomes
Computational prediction of origin of replication (ORI) has been of great interest in bioinformatics and several methods including GC Skew, Z curve, auto-correlation etc. have been explored in the past. In this paper, we…
AIS-MACA- Z: MACA based Clonal Classifier for Splicing Site, Protein Coding and Promoter Region Identification in Eukaryotes
Bioinformatics incorporates information regarding biological data storage, accessing mechanisms and presentation of characteristics within this data. Most of the problems in bioinformatics and be addressed efficiently by…
Lenia and Expanded Universe
We report experimental extensions of Lenia, a continuous cellular automata family capable of producing lifelike self-organizing autonomous patterns. The rule of Lenia was generalized into higher dimensions, multiple kern…
Artificial Life