A New Look at the Classical Entropy of Written English

Computer Science – Computation and Language

Scientific paper

Rate now

  [ 0.00 ] – not rated yet Voters 0   Comments 0

Details

Submitted to the IEEE Transactions on Information Theory. Reference on page 5 corrected. Weighted average of HL on page 12 cor

Scientific paper

A simple method for finding the entropy and redundancy of a reasonable long sample of English text by direct computer processing and from first principles according to Shannon theory is presented. As an example, results on the entropy of the English language have been obtained based on a total of 20.3 million characters of written English, considering symbols from one to five hundred characters in length. Besides a more realistic value of the entropy of English, a new perspective on some classic entropy-related concepts is presented. This method can also be extended to other Latin languages. Some implications for practical applications such as plagiarism-detection software, and the minimum number of words that should be used in social Internet network messaging, are discussed.

No associations

LandOfFree

Say what you really think

Search LandOfFree.com for scientists and scientific papers. Rate them and share your experience with other people.

Rating

A New Look at the Classical Entropy of Written English does not yet have a rating. At this time, there are no reviews or comments for this scientific paper.

If you have personal experience with A New Look at the Classical Entropy of Written English, we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and A New Look at the Classical Entropy of Written English will most certainly appreciate the feedback.

Rate now

     

Profile ID: LFWR-SCP-O-150068

  Search
All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.