Foreground and Background Lexicons and Word Sense Disambiguation for Information Extraction

Computer Science – Computation and Language

Scientific paper

Rate now

  [ 0.00 ] – not rated yet Voters 0   Comments 0

Details

12 pages

Scientific paper

Lexicon acquisition from machine-readable dictionaries and corpora is currently a dynamic field of research, yet it is often not clear how lexical information so acquired can be used, or how it relates to structured meaning representations. In this paper I look at this issue in relation to Information Extraction (hereafter IE), and one subtask for which both lexical and general knowledge are required, Word Sense Disambiguation (WSD). The analysis is based on the widely-used, but little-discussed distinction between an IE system's foreground lexicon, containing the domain's key terms which map onto the database fields of the output formalism, and the background lexicon, containing the remainder of the vocabulary. For the foreground lexicon, human lexicography is required. For the background lexicon, automatic acquisition is appropriate. For the foreground lexicon, WSD will occur as a by-product of finding a coherent semantic interpretation of the input. WSD techniques as discussed in recent literature are suited only to the background lexicon. Once the foreground/background distinction is developed, there is a match between what is possible, given the state of the art in WSD, and what is required, for high-quality IE.

No associations

LandOfFree

Say what you really think

Search LandOfFree.com for scientists and scientific papers. Rate them and share your experience with other people.

Rating

Foreground and Background Lexicons and Word Sense Disambiguation for Information Extraction does not yet have a rating. At this time, there are no reviews or comments for this scientific paper.

If you have personal experience with Foreground and Background Lexicons and Word Sense Disambiguation for Information Extraction, we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and Foreground and Background Lexicons and Word Sense Disambiguation for Information Extraction will most certainly appreciate the feedback.

Rate now

     

Profile ID: LFWR-SCP-O-541336

  Search
All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.