Methods and algorithms for statistical analysis of protein sequences

Brendel, V.; Bucher, P.; Nourbakhsh, I. R.; Blaisdell, B. E.; Karlin, S.

doi:10.1073/pnas.89.6.2002

research article

Methods and algorithms for statistical analysis of protein sequences

Brendel, V.

•

Bucher, P.

•

Nourbakhsh, I. R.

more

1992

Proceedings Of The National Academy Of Sciences Of The United States Of America

We describe several protein sequence statistics designed to evaluate distinctive attributes of residue content and arrangement in primary structure. Considered are global compositional biases, local clustering of different residue types (e.g., charged residues, hydrophobic residues, Ser/Thr), long runs of charged or uncharged residues, periodic patterns, counts and distribution of homooligopeptides, and unusual spacings between particular residue types. The computer program SAPS (statistical analysis of protein sequences) calculates all the statistics for any individual protein sequence input and is available for the UNIX environment through electronic mail on request to V.B. (volker/genomic@stanford.edu).

Type

research article

DOI

10.1073/pnas.89.6.2002

Authors

Brendel, V.

•

Bucher, P.

•

Nourbakhsh, I. R.

•

Blaisdell, B. E.

•

Karlin, S.

Publication date

1992

Published in

Proceedings Of The National Academy Of Sciences Of The United States Of America

Volume

89

Issue

6

Start page

2002

End page

6

Note

Department of Mathematics, Stanford University, CA 94305-2125.

Peer reviewed

REVIEWED

EPFL units

GR-BUCHER

Available on Infoscience

December 17, 2007

Use this identifier to reference this record

https://infoscience.epfl.ch/handle/20.500.14299/15654