Importance Sampling of Word Patterns in DNA and Protein Sequences

Statistics – Applications

Scientific paper

Rate now

  [ 0.00 ] – not rated yet Voters 0   Comments 0

Details

Scientific paper

Monte Carlo methods can provide accurate p-value estimates of word counting test statistics and are easy to implement. They are especially attractive when an asymptotic theory is absent or when either the search sequence or the word pattern is too short for the application of asymptotic formulae. Naive direct Monte Carlo is undesirable for the estimation of small probabilities because the associated rare events of interest are seldom generated. We propose instead efficient importance sampling algorithms that use controlled insertion of the desired word patterns on randomly generated sequences. The implementation is illustrated on word patterns of biological interest: Palindromes and inverted repeats, patterns arising from position specific weight matrices and co-occurrences of pairs of motifs.

No associations

LandOfFree

Say what you really think

Search LandOfFree.com for scientists and scientific papers. Rate them and share your experience with other people.

Rating

Importance Sampling of Word Patterns in DNA and Protein Sequences does not yet have a rating. At this time, there are no reviews or comments for this scientific paper.

If you have personal experience with Importance Sampling of Word Patterns in DNA and Protein Sequences, we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and Importance Sampling of Word Patterns in DNA and Protein Sequences will most certainly appreciate the feedback.

Rate now

     

Profile ID: LFWR-SCP-O-231820

  Search
All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.