Since life began on Earth, the four types of bases (A, G, C, and T(U)) that form two sets of base pairs have remained unchanged as the components of nucleic acids that replicate and transfer genetic information. Throughout evolution, except for the U to T modification, the four base structures have not changed. This constancy within the genetic code raises the question of how these complicated nucleotides were generated from the molecules in a primordial soup on the early Earth. At some prebiotic stage, the complementarity of base pairs might have accelerated the generation and accumulation of nucleotides or oligonucleotides. We have no clues whether one pair of nucleobases initially appeared on the early Earth during this process or a set of two base pairs appeared simultaneously. Recently, researchers have developed new artificial pairs of nucleobases (unnatural base pairs) that function alongside the natural base pairs. Some unnatural base pairs in duplex DNA can be efficiently and faithfully amplified in a polymerase chain reaction (PCR) using thermostable DNA polymerases. The addition of unnatural base pair systems could expand the genetic alphabet of DNA, thus providing a new mechanism for the generation novel biopolymers by the site-specific incorporation of functional components into nucleic acids and proteins. Furthermore, the process of unnatural base pair development might provide clues to the origin of the natural base pairs in a primordial soup on the early Earth. In this Account, we describe the development of three representative types of unnatural base pairs that function as a third pair of nucleobases in PCR and reconsider the origin of the natural nucleic acids. As researchers developing unnatural base pairs, they use repeated "proof of concept" experiments. As researchers design new base pairs, they improve the structures that function in PCR and eliminate those that do not. We expect that this process is similar to the one functioning in the chemical evolution and selection of the natural nucleobases. Interestingly, the initial structures designed by each research group were quite similar to those of the latest successful unnatural base pairs. In this regard, it is tempting to form a hypothesis that the base pairs on the primordial Earth, in which the natural purine bases, A and G, and pyrimidine bases, C and T(U), originated from structurally similar compounds, such as hypoxanthine for a purine base predecessor. Subsequently, the initial base pair evolved to the present two sets of base pairs via a keto-enol tautomerization of the initial compounds.
Read full abstract