Course: MDMBIO1: Introduction to Bioinformatics
Notes of Module 4
Genomics & Human Genome Project
Content:
Genome organization and structure, Sequencing techniques:
Sanger, Next
Generation Sequencing (NGS), Nanopore, Applications: disease
gene identification, forensic genomics, Human Genome Project:
goals, achievements, ethical issues, Comparative genomics
--
1. Genome Organization and Structure
Genomes are the complete set of genetic instructions for an organism.
Their organization and structure vary dramatically between
prokaryotes and eukaryotes, reflecting the differences in their cellular
complexity. The efficient packaging and regulation of this vast amount
of DNA are essential for life.
Prokaryotic Genome Organization
Prokaryotes, such as bacteria and archaea, have a relatively simple
genome structure optimized for rapid replication and gene expression.
• Key Features:
o Single, Circular Chromosome: The main genome is
typically a single, double-stranded, circular DNA molecule.
Unlike eukaryotes, this DNA is not enclosed in a nucleus.
o Nucleoid Region: The chromosome is located in a dense,
irregularly shaped region of the cytoplasm called the
nucleoid. It's not a membrane-bound organelle.
o High Gene Density: Prokaryotic genomes are very
compact. A high percentage of their DNA consists of coding
sequences, with very little
non-coding or "junk" DNA between genes.
o Plasmids: Many prokaryotes also contain small, extra-
chromosomal DNA molecules called plasmids. These are
circular and replicate independently of the main
chromosome. Plasmids often carry genes
that provide a selective advantage, such as antibiotic
resistance.
o Operons: Genes with related functions are often organized
into operons, which are transcribed together as a single
mRNA molecule. This allows for coordinated gene
expression.
• DNA Compaction:
o To fit the large circular chromosome into the small cell,
prokaryotic
DNA is highly compacted through a process called
supercoiling.
o Supercoiling is the overwinding or underwinding of the DNA
helix, which causes it to coil upon itself. This process is
managed by enzymes
called topoisomerases.
o The supercoiled DNA forms a series of loops, which are
anchored by proteins, creating a highly organized and
condensed structure within the nucleoid.
Eukaryotic Genome Organization
Eukaryotes (animals, plants, fungi) have a much more complex genome
structure due to their larger size, linear chromosomes, and the presence
of a nucleus.
• Key Features:
o Multiple, Linear Chromosomes: The genome is divided into
multiple,
linear, double-stranded DNA molecules. The number of
chromosomes is species-specific (humans have 46).
o Nuclear Enclosure: The chromosomes are contained within
a
membrane-bound nucleus, physically separating
transcription from translation.
o Introns and Exons: Eukaryotic genes are often composed of
coding regions (exons) and non-coding, intervening sequences
(introns). The
introns are removed from the RNA transcript during a
process called splicing to produce the final mRNA.
o Extensive Non-coding DNA: A significant portion of the
eukaryotic genome consists of non-coding DNA. This
includes regulatory sequences, repetitive elements,
pseudogenes, and more.
• Chromatin: The DNA-Protein Complex:
o The eukaryotic DNA is wrapped around proteins called
histones to form a complex called chromatin. This is the
primary level of DNA packaging.
o There are four main types of histones (H2A, H2B, H3, and H4)
that form an octamer (a complex of eight histones).
o The DNA wraps around this histone octamer, forming a
nucleosome, which resembles "beads on a string."
o A fifth histone, H1, helps to link the nucleosomes together.
• Levels of Compaction:
o 10-nm Fiber ("Beads on a String"): This is the basic, least
compacted level of chromatin, where DNA is wrapped around
nucleosomes.
o 30-nm Fiber: The nucleosomes are further coiled into a
more condensed helical structure, known as the 30-nm
fiber. This is the predominant form of chromatin in the
interphase nucleus.
o Looped Domains: The 30-nm fiber is organized into larger
looped
domains that are attached to a protein scaffold.
o Metaphase Chromosome: During cell division, the
looped domains are further condensed to form the highly
compact and visible metaphase chromosome. This
allows for the orderly segregation of chromosomes.
Non-coding DNA in Eukaryotes
Unlike prokaryotes, a large majority of the eukaryotic genome does not
code for proteins. This "non-coding DNA" was once considered "junk,"
but is now known to be functionally important.
• Key Types:
o Introns: Non-coding sequences within genes that are
spliced out during RNA processing. They can play a role in
gene regulation and
allow for alternative splicing, which produces multiple
proteins from a single gene.
o Regulatory Elements: Sequences that control gene expression,
such as
promoters (which initiate transcription), enhancers,
and silencers (which increase or decrease transcription).
o Repetitive DNA: These sequences are repeated
many times throughout the genome. They include:
▪ Tandem Repeats: Short sequences arranged end-to-
end (e.g., at centromeres and telomeres).
▪ Transposable Elements (Transposons): "Jumping
genes" that can move from one location to another,
influencing genome structure and evolution.
o Pseudogenes: Non-functional gene copies that have lost
their ability
to be transcribed or translated.
o Non-coding RNA Genes: Sequences that are
transcribed into functional RNA molecules (tRNA, rRNA,
microRNA) that are not translated into proteins.
Comparison: Prokaryotic vs. Eukaryotic Genomes
Introns/
Feature Prokaryotic Genome Exons Generally absent
Nucleoid region
Location (cytoplasm)
Single, circular
Structure chromosome
Gene High (mostly coding
Density DNA)
Eukaryotic Genome
Nucleus
Multiple, linear chromosomes Low
(large amount of non-coding DNA)
Present
Feature Prokaryotic Genome Eukaryotic Genome
Supercoiling with Chromatin (DNA wrapped
DNA associated around
Packaging proteins histones)
Associated Mitochondria and chloroplast
Plasmids DNA
DNA
2. Sequencing Techniques
1. Sanger Sequencing (First-Generation)
Sanger sequencing, also known as the "chain-termination method,"
was the foundational technology developed in the 1970s. It is highly
accurate for sequencing short DNA fragments.
• Principle: The method uses dideoxynucleotides (ddNTPs),
which are modified nucleotides that lack a 3' hydroxyl group. When
a ddNTP is incorporated into a growing DNA strand, it stops the
polymerization process.
• Process:
1. A DNA sample is mixed with a primer, DNA polymerase,
normal nucleotides (dNTPs), and a small amount of
fluorescently labeled ddNTPs (one for each of the four
bases).
2. This reaction produces millions of DNA fragments of varying
lengths, each ending with a specific, color-coded ddNTP.
3. The fragments are separated by size using capillary
electrophoresis. A laser detects the color of the
fluorescence as the fragments pass by, and a computer reads
the sequence from the smallest to the largest fragment.
2. Next-Generation Sequencing (NGS)
Next-Generation Sequencing (NGS), or Massively Parallel
Sequencing, revolutionized the field by enabling the sequencing of
millions of DNA fragments simultaneously. This dramatically reduced the
cost and time required to sequence an entire genome.
• Principle: NGS doesn't use electrophoresis. Instead, it relies on
"sequencing by synthesis," where millions of DNA fragments are
sequenced in parallel on a single chip, called a flow cell.
• Process:
1. The DNA is fragmented, and specialized adapters are
attached to each piece.
2. These fragments are then bound to a flow cell and amplified
to create clusters of identical DNA molecules.
3. Fluorescently labeled nucleotides are added one at a time. As
each base is incorporated, a camera captures its unique
fluorescent signal.
4. The fluorescent tag is then removed, and the process repeats
for the next base. A computer then aligns the millions of
short sequence reads to reconstruct the full DNA sequence.
3. Nanopore Sequencing (Third-Generation)
Nanopore sequencing is a modern technology that sequences
single DNA molecules in real time, without the need for
fragmentation or amplification.
• Principle: This technique uses a nanopore, a tiny protein channel
embedded in a membrane. As a DNA strand passes through the
pore, it disrupts an electrical current flowing across the membrane.
• Process:
1. A single DNA molecule, unzipped into a single strand by
a motor protein, is threaded through the nanopore.
2. Each passing base (A, T, C, G) causes a unique change in
the electrical current.
3. A computer measures these current fluctuations and
translates them into a DNA sequence in real time. This
method can generate exceptionally long reads, which
simplifies the assembly of complex genomes.
3. Applications of Genomics
Genomics, the study of an organism's entire genome, has revolutionized
biology and medicine. By providing a comprehensive view of an organism's
genetic blueprint, it has opened up a wide range of applications that were
previously impossible.
Disease Gene Identification and Personalized Medicine
One of the most significant applications of genomics is in understanding,
diagnosing, and treating human diseases. By sequencing and analyzing
human genomes, scientists can identify the genetic variations that
contribute to disease.
Disease Gene Identification
Genomics has been instrumental in finding the genes responsible for
both rare and common diseases.
• Monogenic Disorders: For diseases caused by a single gene
mutation, like cystic fibrosis or Huntington's disease, genomics
allows for the pinpointing of the exact genetic error. This enables
accurate genetic testing, carrier screening, and preimplantation
genetic diagnosis.
• Complex Diseases: For common diseases like heart disease,
diabetes, and cancer, which are influenced by multiple genes and
environmental factors, genomics uses methods like Genome-Wide
Association Studies (GWAS). GWAS compares the genomes of
large groups of people with a disease to those without it to find
genetic markers that are more common in the affected group. This
helps identify new risk factors and pathways for disease.
Personalized Medicine (Precision Medicine)
Genomics is the foundation of personalized medicine, which
tailors medical treatment to an individual's unique genetic
makeup.
• Pharmacogenomics: This field studies how an individual's
genetic variations affect their response to drugs. By analyzing a
person's genome, doctors can predict whether a drug will be
effective, ineffective, or cause adverse side effects. For example,
genetic testing for certain enzymes can help doctors prescribe the
correct dosage of blood thinners, like warfarin, to prevent
dangerous side effects.
• Targeted Therapies: In cancer treatment, genomics is used to
analyze the DNA of a tumor. This allows for the identification of
specific mutations that drive tumor growth, enabling the use of
targeted therapies that are designed to attack only the cancer
cells with those specific mutations, leading to more effective
treatment with fewer side effects.
Forensic Genomics
Forensic genomics applies genomic techniques to criminal investigations
and human identification. It has become an essential tool for law
enforcement.
• DNA Fingerprinting: The most well-known forensic application is
DNA fingerprinting, which analyzes highly variable regions of the
genome called Short Tandem Repeats (STRs). Because the
number of repeats in these regions is unique to each individual (with
the exception of identical twins), a DNA profile can be created from a
biological sample (e.g., blood, hair, saliva) found at a crime scene.
This profile can then be compared to a suspect's DNA or searched
against national DNA databases like CODIS to identify a match.
• Forensic Genetic Genealogy: This cutting-edge application uses
genomic data to solve cold cases. By analyzing a crime scene DNA
sample, a comprehensive genetic profile is created and uploaded
to public genealogy databases. Investigators can then identify
distant relatives of the suspect and use traditional genealogical
research to build a family tree and narrow down the list of
suspects.
Genomics in Agriculture and Biotechnology
Genomics is transforming agriculture by providing powerful tools to
improve crop and livestock traits.
• Crop Improvement: Genomics helps scientists identify genes
that control desirable traits in crops, such as increased yield,
resistance to pests and diseases, and improved nutritional
value.
o Marker-Assisted Selection (MAS): Instead of waiting for a
plant to grow and exhibit a trait, breeders can use DNA
markers linked to that trait to select the most promising
seedlings, significantly accelerating the breeding process.
o CRISPR Gene Editing: Tools like CRISPR-Cas9 allow for
precise and targeted editing of a plant's genome. This can
be used to create crops with enhanced traits, such as
disease resistance or longer shelf life, without introducing
foreign DNA.
• Livestock Improvement: Similar to crops, genomics is used in
animal husbandry to select animals with superior genetic traits,
leading to healthier and more productive livestock. It also helps
identify genetic markers for disease susceptibility, allowing for
preventative measures and improved animal welfare.
Comparative Genomics and Evolutionary Studies
Comparative genomics is the study of the similarities and differences
between the genomes of different species. This field provides a detailed
view of how organisms are related and how evolution has shaped their
genomes.
• Identifying Conserved Genes: By comparing the genomes of
many species, scientists can identify genes and DNA sequences that
have been conserved (remained unchanged) over millions of years
of evolution. These conserved regions often represent genes that
are essential for life, such as those involved in basic cellular
functions.
• Understanding Evolution: Comparing the genomes of closely
related species, like humans and chimpanzees, reveals the genetic
differences that make each species unique. This helps us
understand the evolutionary changes that led to the development
of specific traits, behaviors, and diseases in different lineages.
• Model Organisms: Comparative genomics allows us to study the
genomes of model organisms (e.g., mice, fruit flies) to gain
insights into human biology and disease. The discovery that many
human genes have a homologous counterpart in these simpler
organisms has made them invaluable for biomedical research.
[Link] Human Genome Project (HGP)
The Human Genome Project (HGP) was a monumental, international
research effort that successfully sequenced and mapped the entire
human genome. Launched in 1990 and completed in 2003, this ambitious
project laid the groundwork for modern genomics, with far-reaching
impacts on medicine, science, and society.
Goals and Methodology
The HGP's primary goals were to create a comprehensive biological
"map" of human DNA.
• Sequencing the Genome: The main objective was to determine
the precise order of the 3.2 billion base pairs that make up the
human genome. The project initially aimed for a 99.9% accuracy
rate and successfully achieved it.
• Gene Identification: Another key goal was to identify and
map all of the estimated 20,000 to 25,000 genes within the
human genome.
• Technological Advancements: A crucial aspect was the
development of new, more efficient, and cost-effective technologies
for sequencing and data analysis.
• Data Accessibility: A core principle was to make all generated
data publicly available, promoting rapid and open scientific
discovery.
• Ethical, Legal, and Social Issues (ELSI): The project dedicated a
portion of its budget to studying the ethical, legal, and social
implications of genomic research, a first for a large-scale scientific
endeavor.
The HGP primarily used a method called hierarchical shotgun
sequencing. This involved:
1. Creating a Physical Map: The genome was first broken into
large, manageable fragments of about 150,000 base pairs. These
fragments were then cloned into Bacterial Artificial
Chromosomes (BACs) and mapped to their original
chromosomal locations.
2. Sanger Sequencing: Each BAC was then broken down into
smaller, overlapping fragments, which were sequenced using the
Sanger sequencing method.
3. Assembly: Powerful computer programs were used to
reassemble the sequenced fragments by identifying
overlapping regions, eventually reconstructing the full
sequence of each chromosome.
Achievements and Scientific Impact
The HGP was completed ahead of its initial 15-year schedule,
providing a foundational reference for all of human genetics.
• The Human Reference Genome: The project delivered a
high-quality, "finished" sequence of the human genome. This
reference genome is a composite of DNA from several
anonymous donors and serves as the standard against which
new genomes are compared.
• Catalyst for Technology: The HGP drove massive innovations in
sequencing technology, leading to the development of automated
Sanger sequencers and laying the groundwork for Next-
Generation Sequencing (NGS). The cost of sequencing a human
genome has since plummeted from billions of dollars to just a few
hundred.
• Model for Collaboration: The HGP was a massive international
collaboration, involving scientists from the United States, the United
Kingdom, France, China, and other countries. This precedent of
large-scale, international "big science" projects has influenced many
subsequent biological and physical science initiatives.
• Rise of Bioinformatics: The sheer volume of data generated
by the HGP necessitated the creation of new computational
tools and databases, effectively giving birth to the field of
bioinformatics.
Ethical, Legal, and Social Implications (ELSI)
The HGP was unique in its proactive approach to the ethical challenges
raised by the new genetic knowledge. The ELSI Research Program was
an integral part of the project.
• Genetic Privacy: The program addressed concerns about who
would have access to an individual's genetic information and how
to prevent its misuse. This led to discussions and legislation
aimed at protecting genetic privacy.
• Genetic Discrimination: A major concern was the potential for
discrimination in employment or health insurance based on an
individual's genetic predisposition to certain diseases. The ELSI
program helped lead to the development of legal protections, such
as the Genetic Information Nondiscrimination Act (GINA) in
the U.S.
• Public Education: The HGP recognized the need to educate the
public on the science of genetics and the project's implications. The
ELSI program funded educational initiatives to foster an informed
public debate on these issues.
• Eugenics and Social Justice: The HGP and its ELSI program
confronted historical fears of eugenics and addressed concerns
about the potential for using genetic information to reinforce social
inequalities or racial biases.
Legacy and Future Directions
The legacy of the HGP extends far beyond the sequencing of a single
genome. It has permanently altered the landscape of biological research
and medicine.
• Personalized Medicine: The HGP is the foundation for
personalized medicine, where medical treatments are tailored
to an individual's genetic makeup. This is particularly
transformative in fields like cancer, where therapies can now
target specific genetic mutations in a tumor.
• New Drug Discovery: With a complete map of human genes,
scientists can now more effectively identify potential drug
targets, leading to faster and more efficient drug discovery and
development.
• Understanding Human Health: The HGP has enabled the
identification of thousands of genes associated with diseases,
from monogenic disorders like cystic fibrosis to complex
conditions like diabetes and heart disease.
• The T2T (Telomere-to-Telomere) Consortium: While the HGP
was declared "complete" in 2003, it left small gaps in the most
repetitive and difficult-to-sequence regions of the genome. In 2022,
the T2T Consortium announced the first truly complete, gapless
sequence of the human genome, a direct result of the
technological foundation laid by the HGP.
[Link] Genomics
Comparative Genomics
Comparative genomics is a field of biological research that compares
the genome sequences of different species. By identifying similarities and
differences, we can gain a deeper understanding of evolutionary
relationships, the function of genes and non-coding DNA, and the genetic
basis of species-specific traits. This field has been a cornerstone of post-
Human Genome Project research, transforming our understanding of
biology.
1. Core Concepts and Principles
The fundamental principle of comparative genomics is that
conservation implies function. DNA sequences that are similar across
many species likely serve a crucial purpose and have been maintained
by natural selection. Conversely, sequences that are highly different or
unique to a single species may be responsible for traits that define that
species.
• Homology: The central concept is homology, which means that
genes or DNA sequences in different species share a common
evolutionary ancestor. There are two types of homologous genes:
o Orthologs: Genes in different species that evolved from a
common ancestral gene through speciation. Orthologs
typically retain the same function in the different species.
For example, the hemoglobin gene in humans and mice are
orthologs.
o Paralogs: Genes within a single species that arose from a
gene duplication event. Paralogs often evolve to have new
functions. For example, the alpha- and beta-globin genes in
humans are paralogs.
• Sequence Conservation: Comparing genomes allows researchers
to identify conserved sequences, which are regions of DNA that
have changed very little over evolutionary time. These regions are
often genes, regulatory elements, or other functional elements.
• Synteny: This refers to the conservation of gene order on a
chromosome. For example, large blocks of genes are often found in
the same order on a human chromosome as on a mouse
chromosome, indicating a shared evolutionary history.
[Link] in Comparative Genomics
Comparative genomics relies on a suite of computational and
experimental methods to analyze and compare genomes.
Multiple Sequence Alignment (MSA)
MSA is the foundational step. It aligns three or more biological
sequences (DNA, RNA, or protein) to find regions of similarity. The
resulting alignment highlights conserved and variable positions. Tools
like Clustal Omega and MUSCLE are commonly used for this.
Genome Assembly and Annotation
Before comparison can begin, each genome must be sequenced,
assembled, and annotated.
• Assembly: Sequencing technologies produce millions of short
DNA fragments, which must be pieced together to reconstruct
the full genome.
• Annotation: After assembly, computational tools are used to identify
and label genes, regulatory elements, and other functional features in
the genome. This is often done by comparing the new genome to well-
annotated genomes.
Phylogenetic Analysis
Comparative genomics provides the data for phylogenetic analysis,
which reconstructs the evolutionary relationships between species. By
comparing homologous sequences from different species, researchers
can build phylogenetic trees that depict the branching patterns of
evolution. This is a powerful tool for understanding the history of life on
Earth.
3. Applications and Impact
Comparative genomics has a wide range of applications,
revolutionizing our understanding of biology, medicine, and
evolution.
Understanding Human Biology and Disease
Comparing the human genome to that of other species, particularly
model organisms like mice, yeast, and fruit flies, has been invaluable
for understanding human biology.
• Gene Function: A gene's function in humans can often be
inferred by studying its ortholog in a simpler model organism. For
example, by studying the function of a homologous gene in a fruit
fly, we can gain insights into a human disease.
• Disease Mechanisms: By comparing human disease-associated
genes with their orthologs in other species, we can create animal
models to study the disease mechanisms and test potential
treatments.
• Non-coding DNA: Comparative genomics has revealed that a
significant portion of our genome's non-coding DNA is highly
conserved and functionally important. These regions are often
regulatory elements that control gene expression.
Evolutionary Biology
Comparative genomics has provided unprecedented insights into the
history of life.
• Reconstructing Evolutionary History: By analyzing genomic
data, scientists can create highly accurate and detailed
phylogenetic trees.
• Identifying Key Evolutionary Innovations: Comparing genomes
of different species can reveal the genetic changes that led to
major evolutionary innovations, such as the development of
multicellularity or the transition from water to land.
• Understanding Adaptation: By comparing the genomes of
closely related species that live in different environments,
researchers can identify the genes responsible for specific
adaptations, such as a polar bear's ability to live in the cold or a
camel's ability to survive in the desert.
Agriculture and Biotechnology
Comparative genomics is transforming agriculture by providing tools
to improve crop and livestock traits.
• Crop Improvement: By comparing the genomes of different crop
varieties, researchers can identify genes linked to desirable traits
like high yield, disease resistance, and drought tolerance. This
information is used in modern breeding programs to develop better
crops.
• Livestock Breeding: Similarly, comparing the genomes of
livestock animals helps in selecting for traits like disease
resistance and improved meat or milk production.
[Link] and Future Directions
Despite its successes, comparative genomics faces several challenges.
• Data Volume and Complexity: As more genomes are
sequenced, the amount of data is growing exponentially.
Analyzing and interpreting this massive dataset requires
increasingly sophisticated computational tools and algorithms.
• Non-Model Organisms: The vast majority of species on Earth have
not had their genomes sequenced. Expanding genomic research to a
wider range of organisms is crucial for a complete understanding of
life's diversity.
• Functional Annotation: While we can identify conserved
sequences, determining their precise function remains a major
challenge. The field of functional genomics is dedicated to
this task.
In the future, comparative genomics will continue to be a driving force
in biology. The development of new sequencing technologies, like
Nanopore sequencing, will enable the sequencing of even more
genomes, providing a richer dataset for comparative analysis. This will
lead to a deeper understanding of everything from human health to the
history of life on Earth.
--
Question Bank:
Part A: 2-Mark Questions
1. What is the primary difference in chromosome structure between
prokaryotic and eukaryotic genomes?
2. Define the term "nucleoid."
3. What is a plasmid, and what is its significance in prokaryotic
genomes?
4. Briefly explain the role of histones in eukaryotic genome
organization.
5. What is the "chain-termination" method in Sanger sequencing?
6. How do dideoxynucleotides (ddNTPs) function to stop DNA
synthesis in Sanger sequencing?
7. What is the main principle behind Next-Generation Sequencing
(NGS)?
8. Define "massively parallel sequencing."
9. What is the key advantage of Nanopore sequencing over other
methods?
10. How does Nanopore sequencing detect the order of
bases in a DNA molecule?
11. What is the primary purpose of forensic genomics?
12. How does DNA fingerprinting work using STRs (Short Tandem
Repeats)?
13. What is the main goal of disease gene identification?
14. What is a Genome-Wide Association Study (GWAS)?
15. State one of the main goals of the Human Genome Project
(HGP).
16. What was one major technological achievement that the HGP
contributed to?
17. What does the acronym ELSI stand for in the context of the
Human Genome Project?
18. Name one ethical concern that was raised by the HGP.
19. Define comparative genomics.
20. What is a homologous gene?
21. What is the difference between an intron and an exon?
22. How is a nucleosome formed?
23. What is supercoiling, and why is it important for prokaryotic
genomes?
24. What is synteny in comparative genomics?
25. How did the HGP contribute to the field of bioinformatics?
Part B: 5-Mark Questions (15 Questions)
26. Describe the difference in gene density between
prokaryotic and eukaryotic genomes and explain why this
difference exists.
27. Outline the key steps in the Sanger sequencing process.
28. Explain the fundamental principles of Next-Generation
Sequencing (NGS) and its main advantage over Sanger sequencing.
29. Describe how Nanopore sequencing works and its
primary benefits for genome assembly.
30. Discuss the application of genomics in disease gene
identification, providing an example of a monogenic disorder.
31. Explain the use of forensic genomics in criminal
investigations, detailing the role of STRs and DNA databases.
32. Describe three major goals of the Human Genome Project.
33. Explain two key achievements of the Human Genome
Project that had a lasting impact on science.
34. Discuss one ethical, legal, and social issue (ELSI) that was
addressed by the Human Genome Project.
35. How does comparative genomics help us to
understand human gene function?
36. Describe the hierarchical levels of DNA packaging in a
eukaryotic chromosome, from the DNA double helix to the
metaphase chromosome.
37. Explain the difference between orthologs and paralogs
and provide an example for each.
38. Describe how a Genome-Wide Association Study (GWAS) is
conducted and what kind of information it provides.
39. Explain why the development of faster, cheaper sequencing
technologies was a critical achievement of the HGP.
40. Discuss the concept of operons and their significance in
prokaryotic gene organization.
Part C: 10-Mark Questions (10 Questions)
41. Compare and contrast the genome organization and
structure of prokaryotes and eukaryotes, discussing chromosomes,
gene density, and DNA packaging.
42. Describe in detail the three major DNA sequencing
techniques (Sanger, NGS, and Nanopore), highlighting their
underlying principles, strengths, and limitations.
43. Discuss the transformative impact of the Human Genome
Project on biomedical research and public health, citing its goals,
key achievements, and ethical considerations.
44. Explain how comparative genomics is used to both understand
human biology and reconstruct evolutionary history.
45. Describe the applications of genomics in both disease gene
identification and forensic genomics, providing detailed examples
for each.
46. Elaborate on the ethical, legal, and social issues (ELSI) that
arose from the Human Genome Project, discussing at least three
key concerns and how they were addressed.
47. Trace the evolution of DNA sequencing technology from
the HGP to the present, explaining how each new generation of
sequencing overcame the limitations of the previous one.
48. Discuss the structure and function of the eukaryotic genome
in detail, including the roles of introns, exons, and different levels of
DNA packaging.
49. Explain the importance of a reference genome and how the
one produced by the HGP serves as a foundation for modern
genomics.
50. Describe how the principles of comparative genomics can be
used to identify functionally important non-coding DNA in the
human genome.