A Winnow-Based Approach to Context-Sensitive Spelling Correction

Computer Science – Learning

Scientific paper

Rate now

  [ 0.00 ] – not rated yet Voters 0   Comments 0

Details

To appear in Machine Learning, Special Issue on Natural Language Learning, 1999. 25 pages

Scientific paper

A large class of machine-learning problems in natural language require the characterization of linguistic context. Two characteristic properties of such problems are that their feature space is of very high dimensionality, and their target concepts refer to only a small subset of the features in the space. Under such conditions, multiplicative weight-update algorithms such as Winnow have been shown to have exceptionally good theoretical properties. We present an algorithm combining variants of Winnow and weighted-majority voting, and apply it to a problem in the aforementioned class: context-sensitive spelling correction. This is the task of fixing spelling errors that happen to result in valid words, such as substituting "to" for "too", "casual" for "causal", etc. We evaluate our algorithm, WinSpell, by comparing it against BaySpell, a statistics-based method representing the state of the art for this task. We find: (1) When run with a full (unpruned) set of features, WinSpell achieves accuracies significantly higher than BaySpell was able to achieve in either the pruned or unpruned condition; (2) When compared with other systems in the literature, WinSpell exhibits the highest performance; (3) The primary reason that WinSpell outperforms BaySpell is that WinSpell learns a better linear separator; (4) When run on a test set drawn from a different corpus than the training set was drawn from, WinSpell is better able than BaySpell to adapt, using a strategy we will present that combines supervised learning on the training set with unsupervised learning on the (noisy) test set.

No associations

LandOfFree

Say what you really think

Search LandOfFree.com for scientists and scientific papers. Rate them and share your experience with other people.

Rating

A Winnow-Based Approach to Context-Sensitive Spelling Correction does not yet have a rating. At this time, there are no reviews or comments for this scientific paper.

If you have personal experience with A Winnow-Based Approach to Context-Sensitive Spelling Correction, we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and A Winnow-Based Approach to Context-Sensitive Spelling Correction will most certainly appreciate the feedback.

Rate now

     

Profile ID: LFWR-SCP-O-541149

  Search
All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.