Resources Contact Us Home
Method and apparatus for inferring the topical content of a document based upon its lexical content without supervision

Image Number 9 for United States Patent #5659766.

An iterative method of determining the topical content of a document using a computer. The processing unit of the computer determines the topical content of documents presented to it in machine readable form using information stored in computer memory. That information includes word-clusters, a lexicon, and association strength values. The processing unit beings by generating an observed feature vector for the document being characterized, which indicates which of the words of the lexicon appear in the document. Afterward, the processing unit makes an initial prediction of the topical content of the document in the form of a topic belief vector. The processing unit uses the topic belief vector and the association strength values to predict which words of the lexicon should appear in the document. This prediction is represented via a predicted feature vector. The predicted feature vector is then compared to the observed feature vector to measure how well the topic belief vector models the topical content of the document. If the topic belief vector adequately model the topical content of the document, then the processing unit's task is complete. On the other hand, if the topic belief vector does not adequately model the topical content of the document, then the processing unit determines how the topic belief vector should be modified to improve the prediction of modeling of the topical content.

  Recently Added Patents
Isolation rings for blocking the interface between package components and the respective molding compound
Reverse mapping method and apparatus for form filling
Modified binding proteins inhibiting the VEGF-A receptor interaction
Avalanche photo diode and method of manufacturing the same
Proton conducting electrolytes with cross-linked copolymer additives for use in fuel cells
Spectral measurement device
  Randomly Featured Patents
Branch pipe for a rotary combustor
Lancet device with skin movement control and ballistic preload
Oscillator for television tuner
Continuous process for manufacturing crystalline zeolites in continuously stirred backmixed crystallizers
Semiconductor diode temperature sensing device
Injection molding machine
Multilayered container
Convergence mechanism for binocular refracting instrument
Composite articles and methods for making the same
Read amplifier for static memories in CMOS technology