Search NASASearch

Engineering topics

Jukes, T. H.

Publications and source records attributed to Jukes, T. H..

At least 19 records

Numerical classification of coding sequences

DNA sequences coding for protein may be represented by counts of nucleotides or codons. A complete reading frame may be abbreviated by its base count, e.g. A76C158G121T74, or with the corresponding codon table, e.g. (AAA)0(AAC)1(AAG)9 ... (TTT)0. We propose that these numerical designations be used to augment current methods of sequence annotation. Because base counts and codon tables do not require revision as knowledge of function evolves, they are well-suited to act as cross-references, for example to identify redundant GenBank entries. These descriptors may be compared, in place of DNA sequences, to extract homologous genes from large databases. This approach permits rapid searching with good selectivity.

Non-NASA Center

Recent evidence for evolution of the genetic code

The genetic code, formerly thought to be frozen, is now known to be in a state of evolution. This was first shown in 1979 by Barrell et al. (G. Barrell, A. T. Bankier, and J. Drouin, Nature [London] 282:189-194, 1979), who found that the universal codons AUA (isoleucine) and UGA (stop) coded for methionine and tryptophan, respectively, in human mitochondria. Subsequent studies have shown that UGA codes for tryptophan in Mycoplasma spp. and in all nonplant mitochondria that have been examined. Universal stop codons UAA and UAG code for glutamine in ciliated protozoa (except Euplotes octacarinatus) and in a green alga, Acetabularia. E. octacarinatus uses UAA for stop and UGA for cysteine. Candida species, which are yeasts, use CUG (leucine) for serine. Other departures from the universal code, all in nonplant mitochondria, are CUN (leucine) for threonine (in yeasts), AAA (lysine) for asparagine (in platyhelminths and echinoderms), UAA (stop) for tyrosine (in planaria), and AGR (arginine) for serine (in several animal orders) and for stop (in vertebrates). We propose that the changes are typically preceded by loss of a codon from all coding sequences in an organism or organelle, often as a result of directional mutation pressure, accompanied by loss of the tRNA that translates the codon. The codon reappears later by conversion of another codon and emergence of a tRNA that translates the reappeared codon with a different assignment. Changes in release factors also contribute to these revised assignments. We also discuss the use of UGA (stop) as a selenocysteine codon and the early history of the code.

Review

Investigations with methanobacteria and with evolution of the genetic code

Mycoplasma capricolum was found by Osawa et al. to use UGA as the code of tryptophan and to contain 75% A + T in its DNA. This change could have been from evolutionary pressure to replace C + G by A + T. Numerous studies have been reported of evolution of proteins as measured by amino acid replacements that are observed when homologus proteins, such as hemoglobins from various vertebrates, are compared. These replacements result from nucleotide substitutions in amino acid codons in the corresponding genes. Simultaneously, silent nucleotide substitutions take place that can be studied when sequences of the genes are compared. These silent evolutionary changes take place mostly in third positions of codons. Two types of nucleotide substitutions are recognized: pyrimidine-pyrimidine and purine-purine interchanges (transitions) and pyriidine-purine interchanges (transversions). Silent transitions are favored when a corresponding transversion would produce an amino acid replacement. Conversely, silent transversions are favored by probability when transitions and transversions will both be silent. Extensive examples of these situations have been found in protein genes, and it is evident that transversions in silent positions predominate in family boxes in most of the examples studied. In associated research a streptomycete from cow manure was found to produce an extracellular enzyme capable of lysing the pseudomurein-contining methanogen Methanobacterium formicicum.

Jukes, T. H.

A change in the genetic code in Mycoplasma capricolum

Mycoplasma capricolum was previously found to use UGA instead of UGG as its codon for tryptophan and to contain 75 percent A + T in its DNA. The codon change could have been due to mutational pressure to replace C + G by A + T, resulting in the replacement of UGA stop codons by UAA, change of the anticodon in tryptophan tRNA from CCA to UCA, and replacement of UGG tryptophan codons by UGA. None of these changes should have been deleterious.

Jukes, T. H.

Evolutionary constraints and the neutral theory

The neutral theory of molecular evolution postulates that nucleotide substitutions inherently take place in DNA as a result of point mutations followed by random genetic drift. In the absence of selective constraints, the substitution rate reaches the maximum value set by the mutation rate. The rate in globin pseudogenes is about 5 x 10 to the -9th substitutions per site per year in mammals. Rates slower than this indicate the presence of constraints imposed by negative (natural) selection, which rejects and discards deleterious mutations.

Jukes, T. H.

Amino acid codes in mitochondria as possible clues to primitive codes

Differences between mitochondrial codes and the universal code indicate that an evolutionary simplification has taken place, rather than a return to a more primitive code. However, these differences make it evident that the universal code is not the only code possible, and therefore earlier codes may have differed markedly from the previous code. The present universal code is probably a 'frozen accident.' The change in CUN codons from leucine to threonine (Neurospora vs. yeast mitochondria) indicates that neutral or near-neutral changes occurred in the corresponding proteins when this code change took place, caused presumably by a mutation in a tRNA gene.

Jukes, T. H.

Recent progress in exobiology and planetary biology

Recent work in the fields of exobiology, the study of the possible characteristics of extraterrestrial life, and planetary biology, the study of life forms as a function of planetary conditions, is reviewed. Searches conducted for life on Mars by the Viking Landers and on Titan by Voyager 1 are considered, and the origin of life on earth is considered in relation to the question of the inorganic trace elements in living systems that are required for life. The question of the origin of terrestrial life from spores carried through the interstellar medium is examined, and the unlikelihood of the survival of such spores except within meteorites or dust particles is pointed out. Studies of organic molecules present in the interstellar medium are indicated as evidence that the conditions necessary for the formation of life can exist in various locations throughout the universe. Investigations of the molecular evolution of life on earth and of life under extreme conditions of heat, cold, drought and ultraviolet radiation, and of the organic compounds found in meteorites and comets are also discussed. The importance of a mechanism of heredity, such as terrestrial DNA, to the evolution of terrestrial and possible extraterrestrial life is pointed out.

Jukes, T. H.

The current status of REH theory

A response is made to the evaluation of Fitch (1980) of REH (random evolutionary hits) theory for the evolutionary divergence of proteins and nucleic acids. Correct calculations for the beta hemoglobin mRNAs of the human, mouse and rabbit in the absence and presence of selective constraints are summarized, and it is shown that the alternative evolutionary analysis of Fitch underestimates the total fixed mutations. It is further shown that the model used by Fitch to test for the completeness of the count of total base substitutions is in fact a variant of REH theory. Considerations of the variance inherent in evolutionary estimations are also presented which show the REH model to produce no more variance than other evolutionary models. In the reply, it is argued that, despite the objections raised, REH theory applied to proteins gives inaccurate estimates of total gene substitutions. It is further contended that REH theory developed for nucleic sequences suffers from problems relating to the frequency of nucleotide substitutions, the identity of the codons accepting silent and amino acid-changing substitutions, and estimate uncertainties.

Holmquist, R.

Neutral changes during divergent evolution of hemoglobins

A comparison of the mRNAs for rabbit and human beta-hemoglobins shows that synonymous changes in codons have accumulated three times as rapidly as nucleotide replacements that produced changes in amino acids. This agrees with predictions based on the so-called neutral theory. In addition, seven codon changes that appear to be single-base changes (according to maximum parsimony) are actually two-base changes. This indicates that the construction of primordial sequences is of limited significance when based on inferences that assume minimum base changes for amino acid replacements.

Jukes, T. H.

Nearest-neighbor doublets in protein-coding regions of MS2 RNA

'Nearest neighbor' base pairs ('doublets') in the protein-coding regions of MS2 RNA have been tabulated with respect to their positions in the first two bases of amino acid codons, in the second two bases, or paired by contact between adjoining codons. Considerable variation is evident between numbers of doublets in each of these three possible positions, but the totals of each of the 16 doublets in the coding regions of the MS2 RNA molecule show much less variation. Compilations of doublets in nucleic acid strands have no predictive value for the amino acid composition of proteins coded by such strands.

Jukes, T. H.

Life on Mars?

Explore the source record for details and available documents.

Jukes, T. H.

On the possible origin and evolution of the genetic code

The genetic code is examined for indications of possible preceding codes that existed during early evolution. Eight of the 20 amino acids are coded by 'quartets' of codons with fourfold degeneracy, and 16 such quartets can exist, so that an earlier code could have provided for 15 or 16 amino acids, rather than 20. If twofold degeneracy is postulated for the first position of the codon, there could have been ten amino acids in the code. It is speculated that these may have been phenylalanine, valine, proline, alanine, histidine, glutamine, glutanic acid, aspartic acid, cysteine and glycine. There is a notable deficiency of arginine in proteins, despite the fact that it has six codons. Simultaneously, there is more lysine in proteins than would be expected from its two codons, if the four bases in mRNA are equiprobable and are arranged randomly. It is speculated that arginine is an 'intruder' into the genetic code, and that it may have displayed another amino acid such as ornithine, or may even have displayed lysine from some of its previous codon assignments. As a result, natural selection has favored lysine against the fact that it has only two codons.

Jukes, T. H.

Recently published protein sequences. III

A listing of polypeptide sequences announced in published papers in 1972 and 1973 is given. The listing contains twenty one new sequencies which were selected by using the distinction of Asn, Asp, Gln, and Glu from Asx and Glx, and the absence of unsequenced peptides as criteria.

Jukes, T. H.