Skip to main content

Protein Structure Database (PDB)


Protein Structure Database (PDB)

Introduction

The Protein Structure Database (PDB) is the primary global repository for the three-dimensional (3D) structures of biological macromolecules such as proteins, nucleic acids, and protein–ligand complexes.
These structures are determined experimentally using techniques like X-ray crystallography, Nuclear Magnetic Resonance (NMR) spectroscopy, and Cryo-Electron Microscopy (Cryo-EM).


PDB plays a vital role in understanding:
Protein structure and function
Molecular interactions
Drug discovery and design
Structural biology and bioinformatics


History and Development
Established in 1971
Founded by Brookhaven National Laboratory (USA)
Initially contained only 7 protein structures
Now maintained by the Worldwide Protein Data Bank (wwPDB)

Members of wwPDB
RCSB PDB (USA)
PDBe (Europe)
PDBj (Japan)
BMRB (Biological Magnetic Resonance Data Bank)


Objectives of PDB
To collect, store, and distribute 3D structural data of biomolecules
To provide free and open access to structural information
To ensure standardization and validation of structural data
To support research and education worldwide

Types of Molecules Stored in PDB
Proteins
Enzymes
Nucleic acids (DNA, RNA)
Protein–protein complexes
Protein–ligand and protein–drug complexes
Virus capsids and ribosomes.


Experimental Methods Used

1. X-ray Crystallography

Most common method
Requires protein crystallization
Provides high-resolution structures

2. NMR Spectroscopy
Used for small proteins
Structure determined in solution
Multiple conformations possible

3. Cryo-Electron Microscopy (Cryo-EM)

Suitable for large complexes
No need for crystallization
Rapidly growing method

PDB File Format

Each structure in PDB is assigned a unique 4-character PDB ID (e.g., 1ABC).

Main Components of a PDB File

HEADER – General information
TITLE – Description of the structure
EXPDTA – Experimental method
ATOM – Atomic coordinates
HETATM – Non-standard atoms (ligands, ions)
SEQRES – Amino acid sequence
CONECT – Bond connectivity
END – End of file


Data Organization in PDB

Primary structure: Amino acid sequence
Secondary structure: α-helices, β-sheets
Tertiary structure: 3D folding
Quaternary structure: Multisubunit assembly


Structural Validation in PDB

Before acceptance, each structure undergoes quality checks:
Resolution assessment
Ramachandran plot analysis
Bond length and angle validation
Steric clash detection
Validation reports are available for each PDB entry.

Access and Tools Provided by PDB

Structure visualization (3D viewers)
Sequence and structure search tools
Structure comparison
Download options in multiple formats
Educational resources.


Applications of PDB

1. Structural Biology

Understanding protein folding
Studying structure–function relationships

2. Drug Discovery

Structure-based drug design
Identification of drug binding sites

3. Bioinformatics

Homology modeling
Molecular docking studies
Protein classification

4. Medical and Clinical Research

Studying disease-related mutations
Designing therapeutic proteins

5. Education

Teaching protein structure and function
Training students in molecular biology


Advantages of PDB

Free and open access
High-quality curated data
Global collaboration
Supports advanced computational analysis


Limitations of PDB

Not all proteins have known structures
Some structures have low resolution
Static structures may not reflect dynamic behavior

Importance of PDB in Modern Biology
PDB acts as a bridge between sequence and function, helping scientists understand how molecular structure determines biological activity. It is an indispensable resource in genomics, proteomics, drug discovery, and biotechnology.


Conclusion

The Protein Structure Database (PDB) is the world’s most important repository for 3D macromolecular structures. By providing reliable and accessible structural data, PDB supports research in biology, medicine, and bioinformatics, making it a cornerstone of modern life sciences.

Comments

Popular Posts

••CLASSIFICATION OF ALGAE - FRITSCH

      MODULE -1       PHYCOLOGY  CLASSIFICATION OF ALGAE - FRITSCH  ❖F.E. Fritsch (1935, 1945) in his book“The Structure and  Reproduction of the Algae”proposed a system of classification of  algae. He treated algae giving rank of division and divided it into 11  classes. His classification of algae is mainly based upon characters of  pigments, flagella and reserve food material.     Classification of Fritsch was based on the following criteria o Pigmentation. o Types of flagella  o Assimilatory products  o Thallus structure  o Method of reproduction          Fritsch divided algae into the following 11 classes  1. Chlorophyceae  2. Xanthophyceae  3. Chrysophyceae  4. Bacillariophyceae  5. Cryptophyceae  6. Dinophyceae  7. Chloromonadineae  8. Euglenineae    9. Phaeophyceae  10. Rhodophyceae  11. Myxophyce...

Mapping of DNA

DNA MAPPING   1. Introduction DNA mapping refers to the process of determining the relative positions of genes or DNA sequences on a chromosome. It provides information about the organization, structure, and distance between genetic markers in a genome. DNA mapping is an essential step toward genome sequencing, gene identification, disease diagnosis, and genetic engineering. DNA maps serve as roadmaps that guide researchers to locate specific genes associated with traits or diseases. 2. Objectives of DNA Mapping To locate genes on chromosomes To determine the order of genes To estimate distances between genes or markers To study genome organization To assist in genome sequencing projects. 3. Principles of DNA Mapping DNA mapping is based on: Recombination frequency Physical distance between DNA fragments Hybridization of complementary DNA Restriction enzyme digestion Use of genetic markers The closer two genes are, the less frequently they recombine during meiosis. 4 . Types of DNA...

Biological Databases – Types of Data and DatabasesNucleotide Sequence Databases (EMBL, GenBank, DDBJ)

Biological Databases – Types of Data and Databases Nucleotide Sequence Databases (EMBL, GenBank, DDBJ) 1. Introduction Biological databases are systematic, computerized collections of biological information that allow efficient storage, retrieval, updating, and analysis of large volumes of biological data. With the advent of genome sequencing, molecular biology, and bioinformatics, biological databases have become essential tools in biological research. These databases support studies in genomics, proteomics, evolutionary biology, taxonomy, medicine, agriculture, and biotechnology. 2. Types of Data Stored in Biological Databases Biological databases store diverse types of biological information, including: 1. Sequence Data DNA sequences RNA sequences Protein sequences 2. Structural Data Three-dimensional structures of proteins Nucleic acid structures 3. Functional Data Gene functions Enzyme activity Regulatory elements 4. Genomic Annotation Data Gene location Exons, introns Promoters a...

Agrobacterium & CaMV-Mediated Gene Transfer –

Agrobacterium and CaMV-Mediated Gene Transfer – Detailed Notes 1. Introduction Gene transfer in plants is often achieved by exploiting natural genetic mechanisms of Agrobacterium tumefaciens and Cauliflower Mosaic Virus (CaMV). These systems allow stable introduction of foreign genes into plant genomes for transgenic plant development. 2. Agrobacterium-Mediated Gene Transfer 2.1 Definition Agrobacterium-mediated gene transfer uses the natural ability of Agrobacterium tumefaciens, a soil bacterium, to transfer a part of its DNA (T-DNA) into plant cells. T-DNA integrates into the plant nuclear genome, enabling stable transformation. 2.2 Mechanism Recognition and attachment Agrobacterium detects phenolic compounds secreted by wounded plant cells. These compounds activate virulence (vir) genes on the Ti (tumor-inducing) plasmid. Activation of vir genes VirA (sensor kinase) and VirG (response regulator) induce expression of other vir genes (VirB, VirC, VirD, VirE). T-DNA processing and tran...

❃HPLC – High Performance Liquid Chromatography

HPLC – High Performance Liquid Chromatography ┏━━━━━ •❃°•°❀°•°❃•━━━━•━━━┓  1. Introduction High Performance Liquid Chromatography (HPLC) is an advanced analytical technique used for the separation, identification, and quantification of components present in a mixture. It is based on the differential distribution of analytes between a stationary phase and a liquid mobile phase under high pressure. HPLC is widely used in biochemistry, biotechnology, pharmaceuticals, food analysis, environmental studies, and clinical diagnostics. 2. Principle of HPLC The principle of HPLC is based on partition, adsorption, ion-exchange, or size-exclusion mechanisms, depending on the type of column used. A liquid mobile phase is pumped at high pressure through a column packed with fine stationary phase particles Sample components interact differently with the stationary phase Components with stronger interaction elute slower Components with weaker interaction elute faster Separated components are detec...

❃HPTLC (HIGH PERFORMANCE THIN LAYER CHROMATOGRAPHY) DETAILED NOTES

HPTLC (HIGH PERFORMANCE THIN LAYER CHROMATOGRAPHY) DETAILED NOTES ┏━━━━━ •❃°•°❀°•°❃•━━━━•━━━┓ 1. INTRODUCTION HPTLC is an advanced form of Thin Layer Chromatography (TLC) that allows high-resolution separation and quantitative analysis of chemical compounds. It combines classical TLC principles with automation, precise sample application, and densitometric detection. HPTLC is widely used in pharmaceuticals, herbal medicine, food analysis, and chemical research. Compared to TLC, HPTLC offers: Better resolution Higher sensitivity Quantitative capabilities Example: Fingerprinting of plant extracts, identification of drugs in mixtures, detection of contaminants in food. 2. PRINCIPLE HPTLC separates compounds based on differential migration on a stationary phase under the influence of a mobile phase. Principle: Adsorption chromatography Compounds interact with the stationary phase (silica gel, alumina, or cellulose) differently depending on polarity, molecular size, or functional groups. Mo...

❃LC-MS (LIQUID CHROMATOGRAPHY – MASS SPECTROMETRY)

LC-MS (LIQUID CHROMATOGRAPHY – MASS SPECTROMETRY)  ┏━━━━━ •❃°•°❀°•°❃•━━━━•━━━┓ 1. INTRODUCTION LC-MS is a hyphenated analytical technique combining Liquid Chromatography (LC) and Mass Spectrometry (MS). It is used for separation, identification, and quantification of compounds in complex mixtures. LC separates analytes based on polarity, size, or charge, while MS detects molecules based on mass-to-charge ratio (m/z). Developed in the 1970s–1980s, LC-MS is now widely used in pharmaceutical, clinical, environmental, and food analysis. Importance : Detects trace levels of compounds (ng–pg range) Analyzes non-volatile, thermally labile compounds that cannot be analyzed by GC-MS Provides structural information through mass fragmentation Example: Detection of drugs in plasma, protein identification in proteomics, pesticide residue analysis in food. 2. COMPONENTS OF LC-MS The LC-MS system has three main parts: A. Liquid Chromatograph (LC) Function: Separates components of a mixture befor...