Gapped BLAST and PSI-BLAST: a new generation of protein database search programs
Data up to Jan 2025
Total Citations Per Year
Abstract
References (82)
Basic local alignment search tool
1990 • 87,358 citations
Random sample consensus
1981 • 23,215 citations
Improved tools for biological sequence comparison.
1988 • 11,455 citations
A general method applicable to the search for similarities in the amino acid sequence of two proteins
1970 • 11,049 citations
Identification of common molecular subsequences
1981 • 9,753 citations
Amino acid substitution matrices from protein blocks.
1992 • 5,998 citations
A Strong Candidate for the Breast and Ovarian Cancer Susceptibility Gene BRCA1
1994 • 5,994 citations
Atlas of protein sequence and structure
1965 • 4,462 citations
Least-Squares Fitting of Two 3-D Point Sets
1987 • 3,657 citations
Sequence Analysis of the Genome of the Unicellular Cyanobacterium Synechocystis sp. Strain PCC6803. II. Sequence Determination of the Entire Genome and Assignment of Potential Protein-coding Regions (Supplement)
1996 • 2,455 citations
Atlas of Protein Sequence and Structure, 1972.
1973 • 2,408 citations
Complete Genome Sequence of the Methanogenic Archaeon, Methanococcus jannaschii
1996 • 2,043 citations
Time Warps, String Edits, and Macromolecules: The Theory and Practice of Sequence Comparison
1983 • 1,947 citations
Detecting Subtle Sequence Signals: a Gibbs Sampling Strategy for Multiple Alignment
1993 • 1,835 citations
An improved algorithm for matching biological sequences
1982 • 1,728 citations
Database of homology‐derived protein structures and the structural meaning of sequence alignment
1991 • 1,641 citations
Rapid similarity searches of nucleic acid and protein data banks.
1983 • 1,577 citations
Methods for assessing the statistical significance of molecular sequence features by using general scoring schemes.
1990 • 1,542 citations
2.2 Mb of contiguous nucleotide sequence from chromosome III of C. elegans
1994 • 1,472 citations
Atlas of Protein Sequence and Structure, 1966.
1967 • 1,463 citations
Profile analysis: detection of distantly related proteins.
1987 • 1,308 citations
Optimal alignments in linear space
1988 • 1,229 citations
The FHIT Gene, Spanning the Chromosome 3p14.2 Fragile Site and Renal Carcinoma–Associated t(3;8) Breakpoint, Is Abnormal in Digestive Tract Cancers
1996 • 1,008 citations
Information content of binding sites on nucleotide sequences
1986 • 953 citations
Issues in searching molecular sequence databases
1994 • 781 citations
Selection of DNA binding sites by regulatory proteins
1987 • 761 citations
A superfamily of conserved domains in DNA damage‐ responsive cell cycle checkpoint proteins
1997 • 750 citations
Statistics of local complexity in amino acid sequences and sequence databases
1993 • 735 citations
Identification of a RING protein that can interact in vivo with the BRCA1 gene product
1996 • 734 citations
[27] Local alignment statistics
1996 • 729 citations
The SWISS-PROT protein sequence data bank and its supplement TrEMBL
1997 • 709 citations
Computer methods to locate signals in nucleic acid sequences
1984 • 673 citations
Amino acid substitution matrices from an information theoretic perspective
1991 • 637 citations
From BRCA1 to RAP1: a widespread BRCT module closely associated with DNA repair
1997 • 525 citations
Identifying protein-binding sites from unaligned DNA fragments.
1989 • 471 citations
Position-based sequence weights
1994 • 422 citations
Applications and statistics for multiple high-scoring segments in molecular sequences.
1993 • 372 citations
Dirichlet mixtures: a method for improved detection of weak but significant protein sequence homology
1996 • 338 citations
A flexible motif search technique based on generalized profiles
1996 • 306 citations
Improved sensitivity of profile searches through the use of sequence weights and gap excision
1994 • 301 citations
A new algorithm for best subsequence alignments with application to tRNA-rRNA comparisons
1987 • 293 citations
Detection of conserved segments in proteins: iterative scanning of sequence databases with alignment blocks.
1994 • 286 citations
Identification of protein sequence homology by consensus template alignment
1986 • 283 citations
Maximum Discrimination Hidden Markov Models of Sequence Consensus
1995 • 248 citations
Prediction of the Coding Sequences of Unidentified Human Genes. VI. The Coding Sequences of 80 New Genes (KIAA0201-KIAA0280) Deduced by Analysis of cDNA Clones from Cell Line KG-1 and Brain
1996 • 230 citations
Optimal sequence alignments
1983 • 226 citations
Complete structure of the hemagglutinin gene from the human influenza A/Victoria/3/75 (H3N2) strain as determined from cloned DNA
1980 • 225 citations
Optimal sequence alignment using affine gap costs
1986 • 222 citations
The statistical distribution of nucleic acid similarities
1985 • 206 citations
Systematic method for the detection of potential λ Cro-like DNA-binding regions in proteins
1987 • 191 citations
Insertional mutagenesis in zebrafish identifies two novel genes, pescadillo and dead eye, essential for embryonic development.
1996 • 184 citations
Weights for data related by a tree
1989 • 182 citations
Volume changes in protein evolution
1994 • 181 citations
Using Dirichlet mixture priors to derive hidden Markov models for protein families.
1993 • 170 citations
GenBank
1997 • 166 citations
Using substitution probabilities to improve position-specific scoring matrices
1996 • 165 citations
Detecting homology of distantly related proteins with consensus sequences
1987 • 155 citations
Limit Distribution of Maximal Non-Aligned Two-Sequence Segmental Score
1994 • 154 citations
Aligning two sequences within a specified diagonal band
1992 • 153 citations
A protein alignment scoring system sensitive at all evolutionary distances
1993 • 151 citations
Distribution of glutamine and asparagine residues and their near neighbors in peptides and proteins.
1991 • 128 citations
A Workbench for large-scale sequence homology analysis
1994 • 121 citations
Maximum-likelihood estimation of the statistical distribution of Smith-Waterman local sequence similarity scores
1992 • 110 citations
The significance of protein sequence similarities
1988 • 100 citations
Weighting aligned protein or nucleic acid sequences to correct for unequal representation
1990 • 95 citations
Embedding strategies for effective use of information from multiple sequence alignments
1997 • 81 citations
Pattern recognition in genetic sequences by mismatch density
1984 • 77 citations
Proceedings of the Fourth International Conference on Intelligent Systems for Molecular Biology
1996 • 76 citations
Proceedings Of The Third International Conference On Intelligent Systems For Molecular Biology
1995 • 75 citations
Analysis of gene duplication repeats in the myosin rod
1983 • 70 citations
A weighting system and aigorithm for aligning many phylogenetically related sequences
1995 • 65 citations
New structure — novel fold?
1997 • 64 citations
Theoretical and Computational Methods in Genome Research
1997 • 59 citations
Recognition of related proteins by iterative template refinement (ITR)
1994 • 46 citations
Locally optimal subalignments using nonlinear similarity functions
1986 • 44 citations
The gal locus from Haemophilus influenzae: cloning, sequencing and the use of gal mutants to study lipopolysaccharide
1992 • 42 citations
The amino acid sequence of leghaemoglobin I from root nodules of broad bean (Vicia faba L.)
1975 • 38 citations
Sequence analysis in the E1 region of adenovirus type 4 DNA
1986 • 32 citations
Locally optimal subalignments using nonlinear similarity functions
1986 • 32 citations
Isolation, characterization, and inactivation of the APA1 gene encoding yeast diadenosine 5',5'''-P1,P4-tetraphosphate phosphorylase
1989 • 31 citations
Rat galactose-1-phosphate uridyltransferase coding sequence, transcription start site and genomic organization
1993 • 9 citations
[Hemoglobins, XXXIII. Note on the Sequence of the hemoglobins of the horse (author's transl)].
1980 • 5 citations
Cited By (0)
No citing papers found in database