Genetic Processes, DNA Replication, Transcription, Translation, and Mendelian Genetics- Week 2
Fundamentals of Nucleic Acids and Amino Acids
Nucleotide Structure: The nucleotide is the fundamental monomeric unit of nucleic acids (DNA and RNA). Each nucleotide is composed of three distinct structural components:
A sugar molecule (deoxyribose in DNA, ribose in RNA).
A phosphate group.
A nitrogenous base.
Nitrogenous Bases and Base Pairing:
Nitrogenous bases form complementary pairs linked by hydrogen bonding across antiparallel strands.
Adenine () pairs specifically with Thymine () in DNA.
Guanine () pairs specifically with Cytosine () in DNA.
In RNA, Uracil () substitutes for Thymine () and pairs with Adenine ().
Strand Geometry and Directionality:
Alternating sugar and phosphate molecules form the covalent structural backbone of nucleic acid strands.
Strands possess defined chemical directionality based on carbon positions on the sugar ring, running from the end to the end.
Complementary DNA strands in a double helix run in opposite directions, referred to as an antiparallel orientation.
Amino Acids:
Amino acids are the monomeric building blocks of polypeptide chains and proteins.
Although numerous amino acids exist in nature, exactly standard amino acids are encoded within human DNA.
Amino acids fall into four distinct chemical classes: basic, nonpolar (hydrophobic), polar (uncharged), and acidic.
Detailed chemical listing of the amino acids encoded in DNA, including chemical names, 3-letter codes, 1-letter codes, molecular formulas, and molecular weights ():
Alanine (, ): Molecular weight (anhydrous ).
Arginine (, ): Basic amino acid; Molecular weight (anhydrous ).
Asparagine (, ): Polar uncharged; Molecular weight (anhydrous ).
Aspartic Acid (, ): Acidic amino acid; Molecular weight (anhydrous ).
Cysteine (, ): Polar uncharged (); Molecular weight (anhydrous ).
Glutamic Acid (, ): Acidic amino acid; Molecular weight (anhydrous ).
Glutamine (, ): Polar uncharged; Molecular weight (anhydrous ).
Glycine (, ): Nonpolar (); Molecular weight (anhydrous ).
Histidine (, ): Basic amino acid; Molecular weight (anhydrous ).
Isoleucine (, ): Nonpolar hydrophobic (); Molecular weight (anhydrous ).
Leucine (, ): Nonpolar hydrophobic (); Molecular weight (anhydrous ).
Lysine (, ): Basic amino acid; Molecular weight (anhydrous ).
Methionine (, ): Nonpolar hydrophobic (); Molecular weight (anhydrous ).
Phenylalanine (, ): Nonpolar hydrophobic; Molecular weight (anhydrous ).
Proline (, ): Nonpolar (); Molecular weight (anhydrous ).
Serine (, ): Polar uncharged (); Molecular weight (anhydrous ).
Threonine (, ): Polar uncharged (); Molecular weight (anhydrous ).
Tryptophan (, ): Nonpolar hydrophobic; Molecular weight (anhydrous ).
Tyrosine (, ): Polar uncharged; Molecular weight (anhydrous ).
Valine (, ): Nonpolar hydrophobic (); Molecular weight (anhydrous ).
Central Dogma and Core Genetic Processes
Central Dogma Framework: Information transfer in biological systems occurs through three central biochemical processes:
Replication: Synthesizing a duplicate copy of DNA from an existing DNA template.
Transcription: Synthesizing an RNA strand from a specific DNA template sequence.
Translation: Decoding mRNA nucleotide sequences into an amino acid chain (protein) at the ribosome.
Mechanism of DNA Replication
Subcellular Localization: DNA replication takes place exclusively inside the cell nucleus in eukaryotes.
Directionality Constraints: DNA polymerases synthesize new DNA strands strictly in the direction.
Semiconservative Replication Model:
Parent DNA double helices unwind by breaking hydrogen bonds between paired nitrogenous bases.
Each separated parental strand serves as a template determining the precise complementary base sequence of a new daughter strand.
Every newly formed daughter DNA double helix consists of one original parental strand and one newly synthesized strand.
Origins of Replication (Sites of Origin):
Replication initiates at specific genomic sequences designated as sites of origin or origins of replication.
Eukaryotic chromosomes contain hundreds to thousands of replication origins across their length.
Replication opens bubbles at origins, which expand laterally in both directions ( and relative to templates) until adjacent bubbles merge.
Transmission electron microscopy (TEM) depicts multiple active replication bubbles along eukaryotic chromosomal DNA (such as in cultured Chinese hamster cells at a scale of ).
Energetics of Synthesis:
Nucleotides enter the synthesis process as nucleoside triphosphates, carrying high-energy phosphate groups.
Hydrolysis of these phosphate groups releases the chemical energy necessary to drive phosphodiester bond formation between nucleotides.
Enzymatic Machinery:
Helicase: Unwinds and separates the double-stranded parental DNA helix at the replication fork.
Single-Strand Binding Proteins (SSBPs): Attach to single-stranded parent DNA to prevent premature re-annealing.
Primase: Lays down a short RNA primer to provide a free group required for DNA polymerase attachment.
DNA Polymerase III: Main elongation enzyme that locks onto primers and adds complementary deoxyribonucleotides sequentially in the direction.
DNA Polymerase I: Removes RNA primers, replaces RNA nucleotides with complementary DNA nucleotides, and conducts initial proofreading.
DNA Ligase: Covalently joins Okazaki fragments together by sealing nicked sugar-phosphate backbones.
Discontinuous Lagging Strand Synthesis:
Because DNA strands run antiparallel and polymerase operates strictly in the direction, assembly differs per strand:
Leading Strand: Synthesized continuously moving toward the advancing replication fork in the direction.
Lagging Strand: Synthesized discontinuously moving away from the replication fork in short segments called Okazaki fragments.
Step-by-Step Lagging Strand Process:
Replication site of origin opens, and primase adds an RNA primer.
DNA Polymerase III synthesizes an Okazaki fragment in the direction away from the fork.
As the fork unwinds further, primase lays down a new primer upstream closer to the fork.
DNA Polymerase III detaches and restarts synthesis from the new primer, generating successive Okazaki fragments.
DNA Polymerase I removes RNA primers and fills the gaps with DNA nucleotides.
DNA Ligase seals the phosphodiester bonds to join adjacent Okazaki fragments into a continuous strand.
DNA Proofreading, Repair Mechanisms, and Mutation Rates
Genomic Scale and Replication Velocity:
Bacterial genome (Escherichia coli): Contains approximately base pairs; complete replication finishes in under .
Human genome: Consists of chromosomes containing approximately base pairs ( base pairs per haploid genome). Copying completes within a few hours.
Printing the human genome single-letter by single-letter (A, C, T, G) would fill over books stacked high.
Fidelity and Accuracy Scale:
Overall final replication error rate is approximately mistake per () bases copied ().
This accuracy is proportionally equivalent to finding specific individual out of the population of Africa, or single user out of everyone on Facebook.
Given base pairs per human haploid genome, an average of nucleotide errors occur per genome replication cycle.
Proofreading and Repair Pathways:
Initial misincorporation rate by DNA polymerase is error per base pairs ( or ).
Proofreading: DNA polymerases feature immediate proofreading activity that detects mispaired bases, exalts incorrect nucleotides, and replaces them during synthesis.
Mismatch Repair: Specialized repair enzymes perform post-replication scanning to excise mispaired bases that escaped initial proofreading.
Nucleotide Excision Repair: Environmental mutagens (e.g., chemicals, radiation) cause structural DNA damage. Nucleases cut out damaged stretches, DNA polymerase synthesizes replacement bases using the undamaged strand as a template, and DNA ligase seals the backbone.
Mutation Frequencies:
Uncorrected mismatches create gene mutations at rates exceeding () per gene.
Due to genome scale, every individual human typically inherits approximately to novel point mutations.
Transcription: Synthesis of RNA from DNA
Definition and Location: Transcription is the DNA-directed synthesis of RNA, occurring inside the nucleus.
Structural Differences: DNA vs. RNA:
Strand Configuration: RNA is single-stranded; DNA is double-stranded.
Length: RNA spans the length of a single gene; DNA spans thousands of genes.
Sugar Unit: RNA contains ribose sugar; DNA contains deoxyribose sugar.
Base Composition: RNA uses Uracil () instead of Thymine (); Uracil pairs with Adenine ().
Functional Gene Architecture:
Promoter: A specific DNA sequence upstream of the gene that acts as the binding site for RNA polymerase and designates the transcription start point.
Coding Region: The segment of DNA transcribed into RNA.
Terminator (Termination Site): A DNA sequence signaling the end of transcription.
Stages of Transcription:
Initiation: RNA polymerase binds to the promoter sequence, unzips local DNA strands, and initiates RNA synthesis at the start point.
Elongation: RNA polymerase advances downstream along the template DNA strand ( template direction), synthesizing a complementary single-stranded RNA transcript in the direction. The DNA double helix rezips behind the advancing enzyme.
Termination: Upon reaching the terminator sequence, RNA polymerase releases the completed primary RNA transcript (pre-mRNA) and detaches from DNA.
Enzymatic Efficiency: Unlike replication, transcription is largely carried out by a single primary enzyme: RNA Polymerase.
Post-Transcriptional Processing and RNA Splicing
Pre-mRNA Processing Modifications:
Before nuclear export, primary pre-mRNA transcripts undergo structural modification, receiving a cap and a poly-A tail (), alongside leader and trailer sequences.
RNA Splicing Mechanics:
Introns (INTRagenic sequences): Non-coding regions within pre-mRNA that are excised and removed.
Exons (EXpressed sequences): Functional coding regions that are retained and joined together to create the mature mRNA transcript.
Memory rule: INtrons are taken OUT; EXons stay IN.
Mature mRNA exits through nuclear pores into the cytoplasm for translation.
Translation and Protein Synthesis
Definition and Location: Translation is the assembly of functional proteins from mature mRNA templates, occurring in the cytoplasm.
Ribosomal Machinery:
Translation is executed by ribosomes, composed of a small ribosomal subunit and a large ribosomal subunit.
Codons and Anticodons:
mRNA genetic information is formatted into nonoverlapping triplets of nitrogenous bases termed codons.
Each codon specifies a single unique amino acid.
Transfer RNA (tRNA) molecules carry an anticodon sequence that binds complementarily to its corresponding mRNA codon.
tRNA delivers its specific amino acid cargo to the active site of the ribosome.
Translation Step-by-Step Sequence:
Mature mRNA binds to the small subunit of the ribosome.
The ribosome identifies the start codon on mRNA.
Specific tRNAs align their anticodons with mRNA codons via complementary base pairing.
The ribosome catalyzes peptide bond formation, linking amino acids into a growing polypeptide chain from the amino (-) end to the acid (-) end.
The ribosome advances along the mRNA in the direction.
Multiple ribosomes can translate a single mRNA simultaneously, forming a polyribosome to rapidly scale protein production.
Protein Structural Hierarchy and Folding
Chaperonin-Assisted Folding:
Newly synthesized linear polypeptide chains must fold into precise three-dimensional shapes to become biologically functional.
Heat shock proteins known as chaperonins assist and direct correct polypeptide folding.
Four Organizational Levels of Protein Structure:
Primary Structure: Linear sequence of amino acids joined together by covalent peptide bonds.
Secondary Structure: Local spatial arrangements formed by hydrogen bonding along the polypeptide backbone, including -helices and -pleated sheets.
Tertiary Structure: Three-dimensional folding pattern of a single polypeptide chain resulting from side-chain (-group) interactions.
Quaternary Structure: Multi-subunit spatial arrangement of two or more distinct polypeptide chains working as a functional complex.
Protein Transport: Folded, mature proteins are transported directly to intracellular or extracellular target sites.
Introduction to Mendelian Genetics and Punnett Squares
Punnett Squares: A grid-based mathematical model used to calculate and predict genotypic and phenotypic ratios in Mendelian inheritance.
Key Inheritance Terminology:
Allele: Alternative nucleotide variant forms of a gene.
Homozygous: Possessing two identical alleles for a specific gene.
Heterozygous: Possessing two different alleles for a specific gene.
Dominant: An allele that expresses its phenotypic trait in both homozygous and heterozygous states, masking recessive alleles.
Recessive: An allele whose phenotypic trait is expressed only in homozygous states.
Codominant: Inheritance pattern where both alleles in a heterozygote are fully and simultaneously expressed.
Sex-Linked: Traits controlled by genes located specifically on sex chromosomes.
Sample Single-Gene Cross (Tongue Rolling):
Example setup: Tongue rolling capability governed by a single gene.
Maternal Genotype: Homozygous dominant for tongue rolling ().
Paternal Genotype: Homozygous recessive for non-tongue rolling ().
Punnett square cross results in heterozygous () offspring expressing the dominant phenotypic trait.
Practice Questions and Self-Assessment
Question 1: What process creates RNA molecules? Give 2 ways RNA is different than DNA?
Answer: The process that creates RNA molecules is Transcription. Three key differences include:
RNA uses Uracil () instead of Thymine ().
RNA is single-stranded, whereas DNA is double-stranded.
RNA uses ribose sugar, whereas DNA uses deoxyribose sugar.
Question 2: These chunks are formed during DNA replication on the lagging strand? By chunking, replication is able to move along the DNA in the correct direction. How do we describe this directionality?
Answer: The discontinuous chunks formed on the lagging strand are Okazaki Fragments. The required directionality of DNA synthesis is strictly in the direction.
Question 3: Every 3 RNA bases code for 1 amino acid. What do we call these triplets?
Answer: These three-base coding sequences are called Codons.
Question 4: Which cellular structure carries out translation to build proteins?
Answer: Translation is carried out by Ribosomes in the cytoplasm.
Question 5: During RNA processing, which sequences are removed, and which stay in?
Answer: Introns are removed (spliced out), while Exons stay in (retained in mature mRNA).