The statistical significance of nucleotide position-weight matrix matches
Generate an AI Snapshot to get a quick, structured summary of this paper.
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
A new version of NMksite, a new version adapted to the processing of nucleotide sequences, is reported, which creates PWM from nucleotide sequence block alignments or occurrence tables using three weight computation schemes.
Abstract
To improve the detection of nucleotide sequence signals (e.g. promoter elements) by position-weight matrices (PWM) using the concept of statistically significant matches. The Mksite program was originally developed for analyzing protein sequences. We report NMksite, a new version adapted to the processing of nucleotide sequences. NMksite creates PWM from nucleotide sequence block alignments or occurrence tables using three weight computation schemes. An original feature of NMksite is the numerical computation of the statistical significance of PWM matches. The utility of this concept is demonstrated in the context of the prediction of splice sites and promoter regions. Mksite and other components of the MODEST (Motif DEsign and Search Tool) package (written in C/Unix) are available at http://igs-server.cnrs-mrs.fr E-mail: [email protected]
