Database searching with DNA and protein sequences: An introduction

descriptionPublicationkeyboard_double_arrow_right Article 01 Jan 2000 English Publisher:Oxford University Press (OUP)Journal:Briefings in Bioinformatics, volume 1, pages 22-32 (issn: 1467-5463, eissn: 1477-4054,

Copyright policy )

Authors: Clare Sansom;

doi: 10.1093/bib/1.1.22

pmid: 11466971

Database searching with DNA and protein sequences: An introduction

- Summary
- Subjects
- Metrics

Abstract

This review of sequence database searching aims to set out current practice in the area, in order to give practical guidelines to the experimental biologist. It describes the basic principles behind the programs and enumerates the range of databases available in the public domain. Of these, the most important are the equivalent DNA databases European Molecular Biology Laboratory (EMBL), GenBank and DNA Databank of Japan (DDBJ), and the protein databases Swiss-Prot and TrEMBL. The commonly used BLAST and FASTA algorithms are described in detail and alternative approaches mentioned briefly. Scoring matrices used to compare amino acid types during protein database searches are compared, with an emphasis on the PAM and BLOSUM series of observed substitution matrices.

Related Organizations

Birkbeck, University of London
United Kingdom
University of London
United Kingdom

Keywords

Genome, Base Sequence, Computational Biology, Information Storage and Retrieval, Proteins, DNA, Europe, Databases as Topic, Japan, Amino Acid Sequence, Algorithms, Software

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	8
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Top 10%
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Top 10%
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

8

Top 10%

Average

bronze

Fields of Science (3) View all

medical and health sciences

basic medicine

Fields of Science

medical and health sciences

basic medicine

View all