Molecular Genetics and Advanced Molecular
Genetics
Comprehensive Study Notes
PART A: MOLECULAR GENETICS (FOUNDATIONS)
1. Introduction to Molecular Genetics
2. Structure of Nucleic Acids
3. DNA Replication
4. Transcription
5. Genetic Code and Translation
6. Regulation of Gene Expression
7. Mutations and DNA Repair
8. Recombination and Genetic Exchange
PART B: ADVANCED MOLECULAR GENETICS
9. Recombinant DNA Technology
10. Polymerase Chain Reaction (PCR) and Variants
11. DNA Sequencing Technologies
12. Genomics and Bioinformatics
13. Genome Editing Technologies
14. Gene Therapy and Molecular Medicine
15. Population and Quantitative Molecular Genetics
16. Emerging Frontiers
17. Quick Revision Summary
PART A: MOLECULAR GENETICS (FOUNDATIONS)
1. Introduction to Molecular Genetics
Molecular genetics is the branch of biology that studies the structure, function, regulation, and
transmission of genetic material at the molecular level. It bridges classical (Mendelian) genetics with
biochemistry and molecular biology, explaining heredity in terms of DNA, RNA, and proteins.
Central Dogma of Molecular Biology (Francis Crick, 1958):
DNA → (Replication) → DNA DNA → (Transcription) → RNA RNA → (Translation) → Protein
Exceptions: Reverse transcription (RNA → DNA, in retroviruses) and RNA replication (in RNA viruses)
show that genetic information can also flow from RNA to DNA/RNA.
2. Structure of Nucleic Acids
2.1 Nucleotide Structure
Each nucleotide consists of: - A pentose sugar (deoxyribose in DNA, ribose in RNA) - A phosphate
group - A nitrogenous base: Purines (Adenine, Guanine) and Pyrimidines (Cytosine, Thymine/Uracil)
2.2 DNA Structure
Watson–Crick double helix model (1953): Two antiparallel strands (5’→3’ and 3’→5’) wound
around a common axis.
Base pairing: A=T (2 hydrogen bonds), G≡C (3 hydrogen bonds) — Chargaff’s rules.
Helix parameters (B-DNA): Right-handed helix, ~3.4 nm per turn, 10 base pairs per turn,
diameter ~2 nm.
DNA forms:
A-DNA: right-handed, compact, dehydrated form
B-DNA: right-handed, most common physiological form
Z-DNA: left-handed, zigzag backbone, occurs in GC-rich regions, may play role in gene
regulation
2.3 RNA Structure
Single-stranded, ribose sugar, uracil replaces thymine.
Types: mRNA (messenger), tRNA (transfer), rRNA (ribosomal), and regulatory/non-coding RNAs
(miRNA, siRNA, lncRNA, snRNA, snoRNA).
2.4 Chromatin and Chromosome Organization
DNA wraps around histone octamers (H2A, H2B, H3, H4) forming nucleosomes (“beads on a
string”).
Higher-order folding: nucleosome → 30 nm fiber → looped domains → condensed chromatin
(heterochromatin) or open chromatin (euchromatin).
Histone modifications: acetylation, methylation, phosphorylation, ubiquitination — form the
“histone code” influencing gene expression.
Linker histone H1 stabilizes higher-order packing.
3. DNA Replication
3.1 General Features
Semiconservative: demonstrated by the Meselson–Stahl experiment (1958) using ¹⁵N-labeled E.
coli DNA.
Occurs during S-phase of the cell cycle.
Bidirectional from origins of replication (oriC in bacteria; multiple origins in eukaryotes).
3.2 Enzymes and Proteins Involved
Enzyme/Protein Function
DnaA Recognizes origin, initiates unwinding
Helicase (DnaB) Unwinds DNA double helix
Single-strand binding (SSB) proteins Stabilize single strands
Topoisomerase/DNA gyrase Relieves supercoiling ahead of fork
Primase (DnaG) Synthesizes RNA primer
DNA polymerase III Main elongation enzyme (prokaryotes)
DNA polymerase I Removes RNA primers, fills gaps (prokaryotes)
DNA ligase Joins Okazaki fragments
DNA polymerase α, δ, ε Eukaryotic replication enzymes
Telomerase Extends telomeres (linear chromosome ends)
3.3 Mechanism
1. Initiation at origin; helicase unwinds DNA forming a replication fork.
2. Primase lays down RNA primers.
3. Leading strand synthesized continuously (5’→3’).
4. Lagging strand synthesized discontinuously as Okazaki fragments.
5. RNA primers removed and replaced with DNA; fragments joined by ligase.
6. Proofreading (3’→5’ exonuclease activity) ensures fidelity (~1 error per 10⁹–10¹⁰ bases).
3.4 Eukaryotic Specifics
Multiple replication origins fire in a coordinated manner (replication licensing via ORC, Cdc6,
Cdt1, MCM helicase complex).
Telomere problem: linear chromosomes cannot fully replicate ends (end-replication problem);
solved by telomerase (ribonucleoprotein enzyme with RNA template) adding TTAGGG repeats.
4. Transcription
4.1 Overview
Synthesis of RNA from a DNA template by RNA polymerase.
4.2 Prokaryotic Transcription
Single RNA polymerase (core enzyme + sigma factor for promoter recognition).
Promoter elements: -10 (Pribnow box, TATAAT) and -35 sequences.
Stages: Initiation → Elongation → Termination (rho-dependent or rho-independent/hairpin-loop
termination).
4.3 Eukaryotic Transcription
Three RNA polymerases:
RNA Pol I — rRNA (28S, 18S, 5.8S)
RNA Pol II — mRNA and most snRNAs (sensitive to α-amanitin)
RNA Pol III — tRNA, 5S rRNA, other small RNAs
Requires general transcription factors (TFIIA, TFIIB, TFIID [binds TATA box via TBP], TFIIE,
TFIIF, TFIIH).
Promoter elements: TATA box, CAAT box, GC box; enhancers/silencers act at a distance via DNA
looping and mediator complex.
4.4 Post-Transcriptional Processing (mRNA maturation, eukaryotes)
5’ capping: 7-methylguanosine cap added co-transcriptionally; aids stability, export, translation
initiation.
3’ polyadenylation: poly-A tail (~200 adenines) added after cleavage at AAUAAA signal.
Splicing: introns removed, exons joined by the spliceosome (snRNPs U1, U2, U4, U5, U6);
recognizes GU-AG splice sites.
Alternative splicing: generates multiple protein isoforms from a single gene, major source of
proteome diversity.
RNA editing: post-transcriptional alteration of nucleotide sequence (e.g., C-to-U editing in ApoB
mRNA).
5. Genetic Code and Translation
5.1 Genetic Code Features
Triplet code: 3 nucleotides = 1 codon.
Degenerate/redundant: multiple codons can specify the same amino acid.
Unambiguous: each codon specifies only one amino acid.
Non-overlapping and comma-less: read continuously in one frame.
Universal (with minor mitochondrial exceptions).
Start codon: AUG (Methionine/fMet). Stop codons: UAA, UAG, UGA.
5.2 Translation Machinery
Ribosomes: prokaryotic 70S (30S + 50S); eukaryotic 80S (40S + 60S).
tRNA: cloverleaf secondary structure, anticodon loop pairs with mRNA codon (wobble hypothesis
explains degeneracy at third codon position).
Aminoacyl-tRNA synthetases: charge tRNAs with correct amino acids (high fidelity
“proofreading”).
5.3 Stages of Translation
1. Initiation: small ribosomal subunit binds mRNA at Shine-Dalgarno sequence (prokaryotes) or 5’
cap (eukaryotes, scanning model); initiator tRNA pairs with AUG; large subunit joins.
2. Elongation: aminoacyl-tRNA enters A site, peptide bond forms (peptidyl transferase, an rRNA-
catalyzed ribozyme activity), translocation shifts ribosome (A→P→E sites).
3. Termination: release factors recognize stop codons, polypeptide released, ribosome dissociates.
5.4 Post-Translational Modifications
Folding (chaperones: Hsp70, chaperonins), cleavage, glycosylation, phosphorylation, ubiquitination,
disulfide bond formation, targeting/sorting signals.
6. Regulation of Gene Expression
6.1 Prokaryotic Regulation — The Operon Model
Lac operon (Jacob & Monod): inducible system controlling lactose metabolism.
Structural genes: lacZ, lacY, lacA
Regulatory: lacI (repressor), operator, promoter
Negative control: repressor blocks transcription in absence of lactose; allolactose inactivates
repressor.
Positive control: CAP-cAMP complex enhances transcription under glucose starvation
(catabolite repression).
Trp operon: repressible system; also regulated by attenuation (premature termination based on
ribosome-RNA polymerase coupling).
6.2 Eukaryotic Gene Regulation Levels
1. Chromatin-level: euchromatin vs heterochromatin, histone modification, nucleosome remodeling
(SWI/SNF complexes).
2. Transcriptional: promoters, enhancers, silencers, insulators, transcription factors, co-
activators/co-repressors.
3. Post-transcriptional: alternative splicing, RNA stability, miRNA-mediated silencing.
4. Translational: initiation factor regulation, mRNA localization.
5. Post-translational: protein modification, degradation (ubiquitin-proteasome system).
6.3 Epigenetic Regulation
DNA methylation: 5-methylcytosine at CpG islands; generally represses transcription.
Histone code: acetylation (activation), methylation (context-dependent), etc.
Genomic imprinting: parent-of-origin-specific gene expression.
X-chromosome inactivation: dosage compensation in mammals via Xist lncRNA.
7. Mutations and DNA Repair
7.1 Types of Mutations
Point mutations: substitution (transition/transversion) — silent, missense, nonsense.
Frameshift mutations: insertions/deletions not in multiples of 3.
Chromosomal mutations: deletion, duplication, inversion, translocation.
Causes: spontaneous (replication errors, tautomeric shifts, depurination/deamination) and induced
(mutagens: chemical — base analogs, alkylating agents, intercalating agents; physical — UV,
ionizing radiation).
7.2 DNA Repair Mechanisms
Mechanism Description
Direct reversal Photoreactivation (photolyase repairs UV-induced
pyrimidine dimers); O6-methylguanine methyltransferase
Base excision repair (BER) Removes damaged single bases via DNA glycosylases
Nucleotide excision repair (NER) Removes bulky lesions (e.g., thymine dimers); defective in
Xeroderma Pigmentosum
Mismatch repair (MMR) Corrects replication errors missed by proofreading;
defective in Lynch syndrome (HNPCC)
Double-strand break repair Homologous recombination (error-free, uses sister
chromatid) and Non-homologous end joining (NHEJ,
error-prone)
SOS response Bacterial emergency repair system, error-prone
translesion synthesis
8. Recombination and Genetic Exchange
Homologous recombination: crossing over during meiosis; Holliday junction model.
Site-specific recombination: e.g., bacteriophage lambda integration.
Transposable elements (“jumping genes”): discovered by Barbara McClintock in maize.
Class I (retrotransposons — via RNA intermediate, e.g., LINEs, SINEs)
Class II (DNA transposons — “cut and paste”, e.g., bacterial IS elements, Tn elements)
Bacterial gene transfer: transformation, conjugation (F factor, Hfr), transduction (generalized
and specialized via bacteriophages).
PART B: ADVANCED MOLECULAR GENETICS
9. Recombinant DNA Technology
9.1 Basic Tools
Restriction endonucleases: bacterial enzymes cutting DNA at specific palindromic sequences
(e.g., EcoRI, HindIII); part of bacterial restriction-modification defense systems.
DNA ligase: joins DNA fragments with compatible ends.
Vectors: plasmids, bacteriophages (lambda), cosmids, BACs (bacterial artificial chromosomes),
YACs (yeast artificial chromosomes) — used to carry foreign DNA into host cells.
Host organisms: E. coli, yeast, mammalian cell lines.
9.2 Steps in Gene Cloning
1. Isolation of gene/DNA fragment of interest.
2. Insertion into vector (using restriction enzymes and ligase).
3. Transformation into host cell.
4. Selection of transformants (antibiotic resistance markers, blue-white screening with lacZ).
5. Screening for the desired clone (via probes, PCR, or sequencing).
9.3 Genomic and cDNA Libraries
Genomic library: represents entire genome, includes introns/regulatory regions.
cDNA library: reverse-transcribed from mRNA using reverse transcriptase; represents only
expressed, spliced sequences.
10. Polymerase Chain Reaction (PCR) and Variants
10.1 Basic PCR
Invented by Kary Mullis (1983). Amplifies specific DNA sequences exponentially using thermostable
DNA polymerase (Taq polymerase).
Cycle steps: 1. Denaturation (~94–96°C): separates DNA strands. 2. Annealing (~50–65°C): primers
bind complementary sequences. 3. Extension (~72°C): polymerase synthesizes new strand.
Repeated for 25–35 cycles, giving exponential amplification (2ⁿ).
10.2 PCR Variants
RT-PCR: reverse transcription PCR, amplifies RNA (via cDNA intermediate).
qPCR/Real-time PCR: quantifies DNA in real time using fluorescent dyes (SYBR Green) or probes
(TaqMan).
Multiplex PCR: multiple primer sets amplify several targets simultaneously.
Nested PCR: increases specificity using two primer sets sequentially.
Digital PCR (dPCR): absolute quantification by partitioning sample into many reactions.
Hot-start PCR, touchdown PCR, allele-specific PCR — specialized variants for
accuracy/specificity.
11. DNA Sequencing Technologies
11.1 First-Generation Sequencing
Sanger (chain-termination) sequencing: uses dideoxynucleotides (ddNTPs) lacking 3’-OH to
terminate chain elongation; fragments separated by capillary electrophoresis; gold standard for
accuracy, low throughput.
Maxam-Gilbert (chemical cleavage) sequencing: largely historical.
11.2 Next-Generation Sequencing (NGS) — Second Generation
Illumina (sequencing by synthesis): reversible dye-terminator chemistry, massively parallel,
high accuracy, dominant platform.
Ion Torrent: detects H⁺ release during nucleotide incorporation (semiconductor sequencing).
454 pyrosequencing: detects pyrophosphate release via chemiluminescence (largely
discontinued).
11.3 Third-Generation (Long-Read) Sequencing
PacBio (SMRT sequencing): real-time single-molecule sequencing, long reads.
Oxford Nanopore: DNA/RNA passes through protein nanopore; changes in ionic current identify
bases; portable, ultra-long reads.
11.4 Applications
Whole genome sequencing (WGS), whole exome sequencing (WES), RNA-seq (transcriptome profiling),
ChIP-seq (protein-DNA interactions), ATAC-seq (chromatin accessibility), single-cell sequencing (scRNA-
seq).
12. Genomics and Bioinformatics
12.1 Key Concepts
Genome: complete set of DNA/genetic material of an organism.
Human Genome Project (1990–2003): mapped ~3.2 billion base pairs of human DNA; ~20,000–
25,000 protein-coding genes.
Comparative genomics: comparing genomes across species to study evolution and function.
Functional genomics: studying gene function and interaction on a genome-wide scale.
Structural genomics: determining 3D protein structures genome-wide.
12.2 Omics Technologies
Field Focus
Genomics DNA/genome sequence
Transcriptomics RNA transcripts (RNA-seq, microarrays)
Proteomics Protein expression (mass spectrometry, 2D gel
electrophoresis)
Epigenomics Epigenetic modifications genome-wide
Metabolomics Small-molecule metabolites
Metagenomics Genetic material from environmental/microbial
communities
12.3 Bioinformatics Tools
Sequence alignment: BLAST, ClustalW/Omega.
Databases: GenBank, EMBL, DDBJ, UniProt, PDB, Ensembl.
Genome browsers: UCSC Genome Browser, Ensembl.
Phylogenetic analysis, molecular docking, and structure prediction (e.g., AlphaFold).
13. Genome Editing Technologies
13.1 CRISPR-Cas9
Derived from bacterial adaptive immune system against phages.
Components: Cas9 endonuclease + guide RNA (gRNA) complementary to target sequence + PAM
(protospacer adjacent motif) recognition.
Cas9 creates a double-strand break; repaired by NHEJ (gene knockout) or homology-directed
repair (HDR, precise editing with a donor template).
Applications: gene knockout/knock-in, functional genomics screens, disease modeling, gene
therapy, agricultural trait improvement.
Newer tools: base editing (direct nucleotide conversion without double-strand breaks), prime
editing (“search-and-replace” genome editing), CRISPR interference (CRISPRi) and activation
(CRISPRa) for gene regulation without cutting DNA.
13.2 Earlier Genome Editing Tools
Zinc Finger Nucleases (ZFNs): engineered zinc-finger DNA-binding domains fused to FokI
nuclease.
TALENs (Transcription Activator-Like Effector Nucleases): TALE DNA-binding domains fused to
FokI; more customizable but more labor-intensive than CRISPR.
13.3 RNA Interference (RNAi)
siRNA/miRNA-mediated gene silencing: double-stranded RNA processed by Dicer into small
fragments, loaded into RISC complex, guides degradation or translational repression of
complementary mRNA.
Used as a research tool for gene knockdown and as a therapeutic modality.
14. Gene Therapy and Molecular Medicine
14.1 Approaches
Viral vectors: adenovirus, adeno-associated virus (AAV), lentivirus — deliver therapeutic genes.
Non-viral methods: liposomes, nanoparticles, electroporation, naked DNA injection.
Ex vivo vs in vivo gene therapy.
CAR-T cell therapy: genetically engineered T-cells expressing chimeric antigen receptors for
cancer immunotherapy.
14.2 Molecular Diagnostics
Southern blotting (DNA), Northern blotting (RNA), Western blotting (protein).
Microarrays: parallel analysis of gene expression/SNP genotyping.
FISH (Fluorescence In Situ Hybridization): detects/localizes specific DNA sequences on
chromosomes.
DNA fingerprinting: uses variable number tandem repeats (VNTRs)/STRs for identity, forensics,
paternity testing.
Molecular markers: RFLP, RAPD, AFLP, SSR/microsatellites, SNPs — used in genetic mapping
and breeding.
15. Population and Quantitative Molecular Genetics
Hardy-Weinberg equilibrium: p² + 2pq + q² = 1; describes allele/genotype frequency stability in
absence of evolutionary forces.
Molecular evolution: neutral theory (Kimura), molecular clocks, synonymous vs non-synonymous
substitution rates (dN/dS ratio).
Linkage disequilibrium: non-random association of alleles at different loci, basis for GWAS
(Genome-Wide Association Studies).
QTL (Quantitative Trait Loci) mapping: identifies genomic regions associated with quantitative
traits.
16. Emerging Frontiers
Synthetic biology: designing and constructing new biological parts, devices, and systems (e.g.,
synthetic genomes, minimal cells).
Epitranscriptomics: chemical modifications of RNA (e.g., m6A methylation) affecting RNA fate.
Single-cell multi-omics: simultaneous profiling of genome, transcriptome, and epigenome at
single-cell resolution.
Optogenetics: light-controlled gene expression/protein activity using light-sensitive proteins.
Liquid biopsy: analysis of circulating tumor DNA/cell-free DNA for non-invasive diagnostics.
Pharmacogenomics: study of how genetic variation affects drug response, guiding personalized
medicine.
Artificial intelligence in genomics: deep learning for variant calling, protein structure
prediction (AlphaFold), and regulatory element prediction.
17. Quick Revision Summary
Central dogma: DNA → RNA → Protein (with reverse transcription as an exception).
Replication is semiconservative, semi-discontinuous, and highly accurate due to proofreading and
repair.
Transcription differs significantly between prokaryotes (single polymerase) and eukaryotes (three
polymerases + extensive RNA processing).
The genetic code is triplet, degenerate, and universal; translation involves initiation, elongation,
and termination.
Gene expression is regulated at chromatin, transcriptional, post-transcriptional, translational, and
post-translational levels.
DNA repair pathways (BER, NER, MMR, recombination repair) safeguard genome integrity; their
failure is linked to cancer and genetic disease.
Recombinant DNA technology, PCR, and sequencing (Sanger → NGS → long-read) underpin modern
molecular genetics research.
CRISPR-Cas9 and related genome-editing tools have revolutionized precision genetic manipulation.
Omics technologies and bioinformatics enable systems-level understanding of genomes.
Applications span medicine (gene therapy, diagnostics, pharmacogenomics), agriculture, forensics,
and evolutionary biology.
End of Notes