Skip to main content

Secondary Databases (PROSITE, PRINTS, BLOCKS)

Secondary Databases (PROSITE, PRINTS, BLOCKS 



Secondary Databases


Introduction

Biological databases are broadly classified into primary and secondary databases.
Primary databases store raw experimental data (e.g., nucleotide or protein sequences), whereas secondary databases contain derived information obtained by analyzing primary sequence data.
Secondary databases are mainly used to:
Identify protein families
Detect conserved motifs, patterns, and domains
Predict protein function
Study structure–function relationships
Examples of secondary databases include PROSITE, PRINTS, BLOCKS, Pfam, etc.


1. PROSITE Database

Definition
PROSITE is a secondary database that documents protein domains, families, and functional sites in the form of patterns and profiles.

Developed by

Swiss Institute of Bioinformatics (SIB)
Maintained along with UniProt
Principle
PROSITE is based on the idea that functionally important regions of proteins are conserved during evolution.
These conserved regions can be represented as:

 1. Patterns (regular expressions)
2. Profiles (position-specific scoring matrices)


Components of PROSITE

Patterns
Short conserved motifs
Written as regular expressions
Useful for identifying active sites or binding sites
Example: Serine protease active site

Profiles
More sensitive than patterns
Can detect distant homologs
Represent the probability of amino acids at each position.

Documentation (PROSITE entries)
Each entry includes:
Description of the protein family/domain
Biological function
References
Links to UniProt


Applications

Protein function prediction
Identification of catalytic and binding sites
Annotation of newly sequenced proteins
Detection of protein families

Advantages
High specificity
Well-curated and annotated
Easy interpretation


Limitations

Patterns may miss distant homologs
False negatives may occur

2. PRINTS Database
Definition
PRINTS is a secondary protein database that identifies protein families using fingerprints, which are groups of conserved motifs.

Developed by

University of Manchester, UK


Principle
Unlike PROSITE, which uses single motifs, PRINTS uses multiple conserved motifs (fingerprints) to characterize a protein family.
A protein is considered a member of a family only if it matches most or all motifs in the fingerprint.


Structure of PRINTS

Each PRINTS entry consists of:
A set of conserved motifs
Alignment of sequences
Functional annotation
Cross-references to other databases


Key Features
Fingerprints improve accuracy
Reduces false positive matches
Useful for family-level classification


Applications

Identification of protein superfamilies
Functional annotation of proteins
Evolutionary studies
Validation of protein family membership

Advantages

High reliability due to multiple motifs
Better discrimination between closely related families
Limitations

Less sensitive to very divergent sequences
Smaller coverage compared to some databases

3. BLOCKS Database
Definition

BLOCKS is a database of conserved regions (blocks) in protein families, represented as ungapped multiple sequence alignments.

Developed by
Fred Hutchinson Cancer Research Center, USA


Principle

A block is a conserved region found in multiple proteins, without insertions or deletions.
These blocks represent functionally or structurally important regions of proteins.


Characteristics
Derived from PROSITE families
Focuses on local conserved regions
Uses position-specific scoring matrices (PSSMs)

BLOCKS Format


Each entry contains:

Protein family name
Conserved block sequences
Alignment information
Scoring matrices


Applications

Detection of conserved motifs
Protein classification
Functional prediction
Sequence similarity searches

Advantages
Highly conserved regions improve accuracy
Ungapped alignments are easy to analyze


Limitations

Ignores variable regions
Limited coverage for novel proteins


Comparison of PROSITE, PRINTS and BLOCKS




Importance of Secondary Databases

Help in functional annotation of proteins
Aid in genome annotation projects
Support comparative genomics and evolutionary studies
Essential tools in bioinformatics and proteomics.


Conclusion

Secondary databases such as PROSITE, PRINTS and BLOCKS play a crucial role in understanding protein structure and function. By analyzing conserved motifs and domains, these databases help in accurate protein classification, functional prediction, and evolutionary analysis, making them indispensable tools in modern bioinformatics.



Comments

Popular Posts

Protein Structure Database (PDB)

Protein Structure Database (PDB) Introduction The Protein Structure Database (PDB) is the primary global repository for the three-dimensional (3D) structures of biological macromolecules such as proteins, nucleic acids, and protein–ligand complexes. These structures are determined experimentally using techniques like X-ray crystallography, Nuclear Magnetic Resonance (NMR) spectroscopy, and Cryo-Electron Microscopy (Cryo-EM). PDB plays a vital role in understanding: Protein structure and function Molecular interactions Drug discovery and design Structural biology and bioinformatics History and Development Established in 1971 Founded by Brookhaven National Laboratory (USA) Initially contained only 7 protein structures Now maintained by the Worldwide Protein Data Bank (wwPDB) Members of wwPDB RCSB PDB (USA) PDBe (Europe) PDBj (Japan) BMRB (Biological Magnetic Resonance Data Bank) Objectives of PDB To collect, store, and distribute 3D structural data of biomolecules To provide free and ope...

DNA FOOTPRINTING

DNA FOOTPRINTING Introduction DNA footprinting is a molecular biology technique used to identify the specific site(s) on DNA where proteins (such as transcription factors) bind. It reveals the exact nucleotide sequences protected by bound proteins against cleavage by nucleases or chemical agents. It is widely used to study DNA-protein interactions, transcription regulation, and gene expression control. Definition DNA footprinting: A technique used to locate the binding site of DNA-binding proteins on DNA by detecting protected regions that are resistant to enzymatic or chemical cleavage. Principle DNA-binding proteins protect the DNA segment they occupy. DNA exposed to nucleases (DNase I) or chemical cleavage agents is cut at accessible regions. Regions bound by protein remain unaffected, leaving a “footprint”. When fragments are separated on a denaturing polyacrylamide gel, the missing bands correspond to protein-binding sites. Key idea: Cleavage occurs everywhere except where the pro...

❥ Southern Blotting Notes

Southern Blotting  ❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥  Introduction Southern blotting is a molecular biology technique used for the detection of specific DNA sequences in a complex mixture of DNA. It was developed by Edwin M. Southern in 1975. The method involves restriction digestion of DNA, separation by gel electrophoresis, transfer (blotting) onto a membrane, and hybridization with a labeled DNA probe. Principle of Southern Blotting The technique is based on the principle of complementary base pairing. A single-stranded labeled DNA probe hybridizes specifically with its complementary DNA sequence immobilized on a membrane. Detection of the label confirms the presence and size of the target DNA fragment. Steps Involved in Southern Blotting. 1. Isolation of DNA Genomic DNA is extracted from cells or tissues. DNA must be pure and intact to ensure accurate results. 2. Restriction Enzyme  Digestion DNA is digested using specific restriction endonucleases. Produces DNA f...

RESTRICTION MAPPING

RESTRICTION MAPPING Introduction Restriction mapping is a molecular biology technique used to determine the relative positions of restriction enzyme recognition sites on a DNA molecule. It involves digestion of DNA with one or more restriction endonucleases followed by analysis of fragment sizes using agarose gel electrophoresis. Restriction mapping is essential for DNA characterization, cloning strategies, gene localization, and genome analysis. Definition Restriction mapping is the process of identifying the number, order, and distances between restriction enzyme cleavage sites within a DNA fragment by analyzing the pattern of fragments generated after enzymatic digestion. Principle Restriction enzymes cut DNA at specific palindromic nucleotide sequences. When DNA is digested with: Single restriction enzyme → produces fragments based on its recognition sites Multiple restriction enzymes → produces fragments whose sizes reveal the relative positions of sites By comparing fragment size...

π“†ž Western Blotting Notes

Western Blotting (Immunoblotting) ❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥ π“†ž❥  Introduction Western blotting, also known as immunoblotting, is a widely used analytical technique for the detection, identification, and quantification of specific proteins in a complex biological sample. The technique combines protein separation by gel electrophoresis with specific antigen–antibody interaction. The method was developed by Towbin et al. (1979) (Burnette 1981---its group work) and is called “Western” in analogy to Southern blotting (DNA) and Northern blotting (RNA). Principle The principle of Western blotting involves: Separation of proteins based on molecular weight using SDS-PAGE Transfer (blotting) of separated proteins onto a membrane Specific detection of the target protein using primary and secondary antibodies Visualization using enzymatic or fluorescent detection systems πŸ‘‰ Antigen–antibody specificity is the core principle of Western blotting. Steps Involved in Western Blotting 1. Sa...

••CLASSIFICATION OF ALGAE - FRITSCH

      MODULE -1       PHYCOLOGY  CLASSIFICATION OF ALGAE - FRITSCH  ❖F.E. Fritsch (1935, 1945) in his book“The Structure and  Reproduction of the Algae”proposed a system of classification of  algae. He treated algae giving rank of division and divided it into 11  classes. His classification of algae is mainly based upon characters of  pigments, flagella and reserve food material.     Classification of Fritsch was based on the following criteria o Pigmentation. o Types of flagella  o Assimilatory products  o Thallus structure  o Method of reproduction          Fritsch divided algae into the following 11 classes  1. Chlorophyceae  2. Xanthophyceae  3. Chrysophyceae  4. Bacillariophyceae  5. Cryptophyceae  6. Dinophyceae  7. Chloromonadineae  8. Euglenineae    9. Phaeophyceae  10. Rhodophyceae  11. Myxophyce...

✩‧₊ Plaque Blotting Technique

Plaque Blotting Technique *ੈ✩‧₊˚༺☆༻*ੈ✩‧₊˚*ੈ✩‧₊˚༺☆༻*ੈ✩‧₊˚ Introduction Plaque blotting is a molecular biology screening technique used to identify specific DNA or RNA sequences present in bacteriophage plaques formed on a bacterial lawn. It is especially useful in the screening of recombinant phage libraries such as Ξ» (lambda) phage genomic or cDNA libraries. This technique combines: Plaque assay (to isolate individual phage clones) Blotting technique (to transfer nucleic acids onto a membrane) Hybridization (to detect specific sequences using labeled probes) Principle of Plaque Blotting The principle of plaque blotting is based on nucleic acid hybridization. Each plaque represents a clone of phage particles containing identical DNA. DNA from phage particles in plaques is: Released Denatured into single strands Transferred onto a nitrocellulose or nylon membrane The membrane is incubated with a labeled DNA/RNA probe complementary to the target sequence. Hybridization between probe and t...

Biological Databases – Types of Data and DatabasesNucleotide Sequence Databases (EMBL, GenBank, DDBJ)

Biological Databases – Types of Data and Databases Nucleotide Sequence Databases (EMBL, GenBank, DDBJ) 1. Introduction Biological databases are systematic, computerized collections of biological information that allow efficient storage, retrieval, updating, and analysis of large volumes of biological data. With the advent of genome sequencing, molecular biology, and bioinformatics, biological databases have become essential tools in biological research. These databases support studies in genomics, proteomics, evolutionary biology, taxonomy, medicine, agriculture, and biotechnology. 2. Types of Data Stored in Biological Databases Biological databases store diverse types of biological information, including: 1. Sequence Data DNA sequences RNA sequences Protein sequences 2. Structural Data Three-dimensional structures of proteins Nucleic acid structures 3. Functional Data Gene functions Enzyme activity Regulatory elements 4. Genomic Annotation Data Gene location Exons, introns Promoters a...

𓆉 INDEX PAGE -NOTETHEPOINT43

INDEX PAGE   MAIN    CONTENT 1.   HSST BOTANY SYLLABUS, DETAILED NOTES, MCQ 2.  SET GENERAL PAPER SYLLABUS, DETAILED NOTES, 50MCQ 3.  SET BOTANY SYLLABUS, DETAILED NOTES, MCQ 4. MSC BOTANY THIRD SEMESTER SYLLABUS, NOTES (KERALA UNIVERSITY ) 5. MSC BOTANY THIRD SEMESTER QUESTION PAPER (KERALA UNIVERSITY ) 6. MSC BOTANY FOURTH SEMESTER SYLLABUS &NOTES (KERALA UNIVERSITY ) 7. FOURTH SEMESTER MSC BOTANY PREVIOUS QUESTION PAPER  (KERALA UNIVERSITY )