There has been a lot of back and forth on Duons. We should come to con consensus before it goes into the main article.
| This is an archive of past discussions about Genetic code. Do not edit the contents of this page. If you wish to start a new discussion or revive an old one, please do so on the current talk page. |
| Archive 1 | Archive 2 |
There has been a lot of back and forth on Duons. We should come to con consensus before it goes into the main article.
My take on the whole this is summed up quite nicely in this article[1] basically the "duon" functionality is already well know and already has defined names like Regulatory DNA sequences, promoters, enhancers, termination sequences. These terms are already used in the Article and I don't think we should be using a term that basically boils down to a PR buzz word. it should not be in the introduction of the article, if at all. Ryftstarr (talk) 14:29, 16 December 2013 (UTC)
This whole Duon stuff is obviously PR buzz and trying to make the fact that epigenetic modifications occur within protein sequences a novel finding AND a distinct phrase is quite laughable. Also, the even larger claim that non-coding selection on protein regions being unknown is even more unbelievable. For transcription factor binding sites within proteins check back to at LEAST 2001, http://nar.oxfordjournals.org/content/29/19/4070.long . This concept is not at all controversial for people in the specific field of epigenetics (modifications to DNA that do not change the underlying genetic code). Also ideas about optimal codons have been expressed since at least 1987 (http://nar.oxfordjournals.org/content/15/3/1281).
Neglecting this background makes the recent insertion both shortsighted and also suspect in seeming to increase the tout of its scientific claims.
REGARDLESS, none of this discussion belongs in an introduction to the genetic code. At best it should be a distant footnote at the end linked to the more extensive discussion on codon usage. Were the editors not aware of this 30+ year research topic (http://en.wikipedia.org/wiki/Codon_usage_bias)?
To be more specific, the last paragraph of the introduction is distracting to a general introduction of the genetic code, which is specifically about the translation of mRNA into a protein sequence. Trying to shoe-horn into some talk about how organisms are more than protein (I happen to agree with things being more than proteins) is obviously out of place and seems like proselytizing. Again, if you really want it to be there, mention it in the context of the main body, not 1/3 of the introduction.
99.174.80.45 (talk) 04:45, 17 December 2013 (UTC)Thomas
I agree with both of those points DMacks. My main dispute was 1) overemphasis to a secondary point in the intro and 2) ignoring the rest of the extensive research on non-coding regions. I agree that the idea of DNA sequence=deterministic is misleading and worth mentioning, but as currently stated it seems to undercut the whole premise rather than the reality being more nuanced. A single sentence linking out to regulatory sequences and possibly codon usage seems like a good idea.
Though, I have to mention the overall misconception about the role most of these regulatory processes play. Most epigenetic modifications, be they transcription factors, DNA methylation, or histone modification, largely relate to changes in how MUCH of a given protein is produced, rather than WHAT protein is produced. There are some recent studies showing that histone/methylation can affect whether introns/exons are included/excluded thereby changing the protein sequence, but that does not (currently) seem to be a primary function. So again, the actual research is much less clear than is currently purported. As it stands, the last paragraph of the intro speaks more about the relevance of protein abundance evolution vs protein sequence evolution. I'm forgetting the more eloquent description of this debate, but it is a long-going discussion in the evolution literature. 99.174.80.45 (talk) 05:15, 17 December 2013 (UTC) Thomas
Here is a rough draft of a replacement sentence: "While the genetic code determines the protein sequence for a given coding region, other genomic regions can influence when and where these proteins are produced" This sentence could be expanded to talk about further impact towards phenotype, but then in my opinion it starts to get bogged down in specifics that are tangential to the main article.99.174.80.45 (talk) 05:04, 18 December 2013 (UTC) Thomas
Genetic code graphic figure GeneticCode21-version-2.svg is confusing in this context. I may be confused myself (not a biologist), but this figure is from the catalog of a company that specializes in posttranslational modification, and makes heavy reference to various modifications, which as far as I can tell have no direct relation to the natural genetic code. My initial interpretation was "oh, so the redundant codons actually specify posttranslational modifications." I can understand the desire for a sexy graphic instead of boring tables, but IMO this page would be improved by simple deletion of that figure.
Robertmacl (talk) 12:45, 13 May 2014 (UTC)
Descriptions of the genetic code have improved in the past ten years, but even a simple definition is still lacking. The old definition is no longer explicitly given - the genetic code is a transfer of linear information from DNA to protein - but it is still strongly implied in everything being said here. The net effect is that there is no working definition for molecular information, and molecular information is the purpose of genetic translations.
I think there is a simple, logical foundation for the genetic code. I think that the genetic code, if it is properly understood, is central to all processes in life. I think there is a phenomenal amount of molecular information stored in and translated by the genetic code, not just codons and amino acids.
I seem to be the only person on the planet that feels this way about it, and that's okay. But I'm a little bit surprised that after ten years these valid ideas are not even mentioned on a page like this. I think for the sake of debate, you should point them out if only to refute them, or tell people why they should reject them.
If anybody cares to understand this, they can start here: http://www.codefun.com/
I absolutely do not mean "genome." If you want to define the genetic code to be "essentially what a codon table shows" then I think that should be included as the first line in the page. Then I think you should explain exactly what a codon table is and exactly what it shows, because other than being defined that way, that is not what the genetic code is.
The basic problem is that "everybody knows" that the genetic code is something that translates "molecular information." Unfortunately, this represents nothing but a tautology in that molecular information is defined as that thing translated by the genetic code, and the genetic code is defined as essentially what a codon table shows.
My basic point is that a codon table is a very small part of what the genetic code actually is, and I am limiting this here specifically to the molecular information translated from nucleotide sequences to protein sequences. The genetic code at that level is still so many things that I think it is incumbent on any explanation like this to clearly define what it is explaining. Short of that, it does more to confuse people than actually clear things up.
This fact is covered by the article but rather much hidden away. It is not represented in the lead. There are two critical steps. The FIRST step is the coupling of the Transfer RNA the the Amino Acid. This requires a specific enzyme, the amino asyl transfer RNA synthetase, for each amino acid. One can say the the DNA code for the AATRS embodies half of the genetic code. But this is not pointed out by the text. It has to be reasoned from the text. --Ettrig (talk) 12:59, 25 November 2014 (UTC)
I would say that tRNA is the molecule that does the translation from codon to amino acid. It is like a dictionary or something. You are correct, it is only because of the tRNA that AAA means Lysine. The translation from AAA to lysine has nothing to do with ribosomes. It is the tRNA and only the tRNA that translates codon to amino acid. "All" that the ribosome does is to get the correct tRNA to match the mRNA and then join the amino acids into a polypeptide. Note that the tRNA also provides the energy for the ribosome to move the mRNA by 3 bases as the mRNA is read. This probably should be clarified -- Lehasa (talk) 13:51, 22 February 2015 (UTC)
Trying to use this data, I found it confusing. It's clearly not percent, since it sums to more than 100. I looked at the column heading, but was not familiar with the percent-like symbol. I hovered over the symbol, and it said "per mille", so I though it was per thousand, but was not sure in what language (I don't know Latin).
I went to the original reference cited in the section, which said "per thousand", which made sense. So I looked up per mille and found that I was not alone is not being familiar with this:
The term occurs so rarely in English that major dictionaries do not agree on the spelling or pronunciation even within a single dialect of English[10] and some major dictionaries such as Macmillan[11] and Longman[12] do not even contain an entry.
So I changed it, so now when you hover over the symbol it says "per thousand", which will be more helpful to the reader, I think. Other opinions are welcome. LouScheffer (talk) 12:29, 17 October 2016 (UTC)
I'm wondering whether the amount of codons per gene varies, and if so, whether there is a minimum and maximum amount of codons per gene. Also, if there's a minimum/maximum amount of codons, is this amount a multiplication of 3 (i.e. 1³, 2³, 3³, ...). That way, we could also know the amount of possible genetic code variations per gene. KVDP (talk) 16:10, 22 June 2017 (UTC)
I was wondering whether there has been any research in Decipherment of the DNA. For instance, there are various types of mutations of the same gene in the human population, which express themselves as differences in real life between the humans.[1]
Logically, each of these mutations is a code for a different message that conveys details on how to do something in the human body. My guess is that each of the 64 codons is a base building block in that code (so comparable to a letter in our own alphabet). Each gene (or hence sequence of codons) will (I think) convey a message to what type of tissue needs to be build (i.e. fat, bone, flesh, ...) and how long this strand of tissue needs to be, and its shape, and to what tissue it should connect). The thickness of the tissue is probably not specified directly, but rather specified by a seperate gene, perhaps via the "codon for specifying length". The latter, I assume because a disease like Talk:Sclerosteosis also exists.
The reason why this is useful to know is because, at present, for treating genetic diseases, we can only just use the genetic code of humans without that disease to overwrite the faulty gene in a person with the disease. However, as Stephen Friend from The Resilience Project found out, there are many versions of "good genetic code", and not all version will work on that person. We don't know why this is, and so every gene therapy that would be undertaken becomes a puzzle, and each gene therapy may need to be repeated several times. If we understand what message is in the gene, we might avoid all this.
KVDP (talk) 09:13, 27 June 2017 (UTC)
References
It seems rather redundant to have both - I undestand the reasons for setting up the table both ways but I don't think it adds much to the article to include the 2nd table. If there are no objection, I'll remove it. Hichris 18:49, 28 November 2006 (UTC)
It's not really redundant. For me (and hopefully for others) this table is a valuable resource that may be used for designing mutagenesis primers when exchanging amino acids by PCR. May I ask you to put it back, please? This message is encrypted! You'll need a brain to decode it. 14:53, 12 January 2007 (UTC)
Here's an alternative presentation, using the IUPAC abbreviations from DNA_sequence:
| Ala | GCN | Leu | YUR, CUN |
|---|---|---|---|
| Arg | CGN, AGR (MGR) | Lys | AAR |
| Asn | AAY | Met | AUG |
| Asp | GAY | Phe | UUY |
| Cys | UGY | Pro | CCN |
| Gln | CAR | Ser | UCN, AGY |
| Glu | GAR | Thr | ACN |
| Gly | GGN | Trp | UGG |
| His | CAY | Tyr | UAY |
| Ile | AUY, AUA (AUH) | Val | GUN |
| START | AUG | STOP | UAR, URA |
Currently, in section "RNA codon table", the header on the "Inverse table for the standard genetic code" table refers to "DNA codons", which should be "RNA codons". It looks like the same template is being used for both DNA and RNA codons, and the substitution T->U is made. However, the column name should be specific. Probably having separate tables would simplify things :) DeepCurl (talk) 16:42, 24 March 2019 (UTC)
As described in my comment on Talk:Proteinogenic_amino_acid#Hydropathy_of_tyrosine, the table in this article in classifying tyrosine as polar and not hydrophobic is inconsistent with other statements in Wikipedia. Tyrosine's own article clarifies that it is near the borderline but "usually classified as" hydrophobic. This article's table should either be changed to match the sourced statements elsewhere, or itself sourced. 2607:FEA8:12A0:44D:0:0:0:C319 (talk) 02:08, 24 May 2020 (UTC)
Take a look at the new 3-D image and consider what is said in the article: "The reason may be that charge reversal (from a positive to a negative charge or vice versa) can only occur upon mutations in the first position, but never upon changes in the second position of a codon." Consider positive {R,K} <-> negative {D,E}. This statement is misleading. Please check me. Charles Juvon (talk) 22:41, 9 September 2020 (UTC)

Both references to the 1.5 x 10^84 number are dead. One possible fix is to insert the actual equation: N[Abs[Sum[(-1)^j*Binomial[21,j]*j^64,{j,21}]],10] = 1.510109516 x 10^84 That's in Mathematica syntax. I'm no good in wiki markup for algebraic equations. N[_,10] is simply formatting. Charles Juvon (talk) 19:11, 15 September 2020 (UTC)
Codon redirects here, but this is not very useful if you more or less know what the genetic code is and are wondering what the heck a codon is. I mean, is it a real physical structure, or is it just a scientific convention? in the first paragraph you get the idea that its a physical structure, later on you learn it can be read from any of three ways. If you chop a strand and have no start/stop sequence do its codons cease to exist? It could use its own article, even if its a short one.Brallan 17:59, 27 March 2007 (UTC)
This article should contain a reference to what ACGTU stand for. There's no obvious way to find the definitions of the symbols if you don't already know the basics. — Preceding unsigned comment added by 71.182.155.221 (talk) 18:18, 9 February 2021 (UTC)
May I suggest that this image would be useful for our readers: [[1]]

Charles Juvon (talk) 22:21, 27 August 2020 (UTC)
H2mex (talk) 03:03, 27 June 2021 (UTC)
Given the current last sentence of the Article and references 99 and 100, we might want to use this material from https://arxiv.org/ftp/arxiv/papers/1303/1303.6739.pdf : "Recent biotech achievements make it possible to employ genomic DNA as data storage more durable than any media currently used (Bancroft et al., 2001; Yachie et al., 2008; Ailenberg & Rotstein, 2009). Perhaps the most direct application for that was proposed even before the advent of synthetic biology. Considering alternative informational channels for SETI, Marx (1979) noted that genomes of living cells may provide a good instance for that. He also noted that even more durable is the genetic code. Exposed to strong negative selection, the code stays unchanged for billions of years, except for rare cases of minor variations (Knight et al., 2001) and context-dependent expansions (Yuan et al., 2010)." ----Charles Juvon (talk) 20:18, 5 November 2020 (UTC)
Unless someone has a better idea, we need to go back to the March 15, 2021 version. The top figure is an embarrassment. That editor is now in red letters. Charles Juvon (talk) 01:45, 17 June 2021 (UTC)
"Degeneracy is a salient feature of genetic codes, because there are more codons than amino acids. The conventional table for genetic codes suffers from an inability of illustrating a symmetrical nature among genetic base codes. In fact, because the conventional wisdom avoids the question, there is little agreement as to whether the symmetrical nature actually even exists. A better understanding of symmetry and an appreciation for its essential role in the genetic code formation can improve our understanding of nature’s coding processes. Thus, it is worth formulating a new integrated symmetrical table for genetic codes, which is presented in this paper. It could be very useful to understand the Nobel laureate Crick’s wobble hypothesis — how one transfer ribonucleic acid can recognize two or more synonymous codons, which is an unsolved fundamental question in biological science."
H2mex (talk) 19:25, 3 July 2021 (UTC)" Charles Juvon (talk) 13:44, 12 July 2021 (UTC)
User 103.172.73.22 changed the image description at the top to say a codon is two nucleotides rather than three. Not clear why they would do that but it is clearly incorrect per other content already on the page. ArbitraryConstant (talk) 21:29, 15 February 2023 (UTC)
Please consider this CC4.0 and use as you wish.
https://www.youtube.com/watch?v=eHZxMAZTFcY Doug youvan (talk) 17:13, 9 September 2023 (UTC)
If another editor feels this would work in the article, please use it. Graph Construction: In our graph, vertices represent the 20 amino acids and the "Stop" signal. An edge connects two vertices if the amino acids they represent can be interchanged through a single point mutation in their corresponding codons. This graph is not just a visualization but an analytical tool, spotlighting the possible amino acid replacements due to minor genetic variations. Highlighting Mechanism: Using the computational capabilities of Mathematica, and with the expertise provided by Centaur Intelligence, each amino acid (and the Stop signal) is successively emphasized. When highlighted, all directly reachable amino acids through a single point mutation are illuminated, thus displaying the mutation landscape for each amino acid. Results: The resultant graph unravels the dense web of interconnections among amino acids based on single point mutations. As we animate through each amino acid, patterns emerge, revealing which amino acids can easily mutate into others and which remain more isolated.
https://www.youtube.com/watch?v=WsGw5w6tiyE
Doug youvan (talk) 01:58, 29 September 2023 (UTC)
It's more accurate, but it would need some work by a graphic artist for scaling. Serine (S) is properly represented. CC 4.0. https://www.researchgate.net/publication/374911250_Amino_Acids_Are_Segregated_in_Hydropathy_-_Molar_Volume_Space_by_the_Second_Position_of_the_Codon Doug youvan (talk) 01:27, 23 October 2023 (UTC)
https://www.researchgate.net/publication/374973672_Codon_Cluster_Analysis_With_Hydropathy_Written_by_GPT-4_in_Python Doug youvan (talk) 19:38, 25 October 2023 (UTC)
This article has been revised as part of a large-scale clean-up project of multiple article copyright infringement. (See the investigation subpage.) Prior content in this article duplicated one or more previously published sources. The material was copied from: https://www.sciencedirect.com/science/article/pii/S1369527405001657. Copied or closely paraphrased material has been rewritten or removed and must not be restored, unless it is duly released under a compatible license. (For more information, please see "using copyrighted works from others" if you are not the copyright holder of this material, or "donating copyrighted materials" if you are.)
For legal reasons, we cannot accept copyrighted text or images borrowed from other web sites or published material; such additions will be deleted. Contributors may use copyrighted publications as a source of information, and, if allowed under fair use, may copy sentences and phrases, provided they are included in quotation marks and referenced properly. The material may also be rewritten, provided it does not infringe on the copyright of the original or plagiarize from that source. Therefore, such paraphrased portions must provide their source. Please see our guideline on non-free text for how to properly implement limited quotations of copyrighted text. Wikipedia takes copyright violations very seriously, and persistent violators will be blocked from editing. While we appreciate contributions, we must require all contributors to understand and comply with these policies. Thank you. DanCherek (talk) 17:00, 15 March 2024 (UTC)
Informasi ini disarikan dari Wikipedia dan disajikan kembali untuk tujuan edukasi. Konten tersedia di bawah lisensi CC BY-SA 3.0. Kami tidak bertanggung jawab atas ketidakakuratan data yang bersumber dari kontribusi publik tersebut.