Resources Contact Us Home
Method and apparatus for inferring the topical content of a document based upon its lexical content without supervision

Image Number 9 for United States Patent #5659766.

An iterative method of determining the topical content of a document using a computer. The processing unit of the computer determines the topical content of documents presented to it in machine readable form using information stored in computer memory. That information includes word-clusters, a lexicon, and association strength values. The processing unit beings by generating an observed feature vector for the document being characterized, which indicates which of the words of the lexicon appear in the document. Afterward, the processing unit makes an initial prediction of the topical content of the document in the form of a topic belief vector. The processing unit uses the topic belief vector and the association strength values to predict which words of the lexicon should appear in the document. This prediction is represented via a predicted feature vector. The predicted feature vector is then compared to the observed feature vector to measure how well the topic belief vector models the topical content of the document. If the topic belief vector adequately model the topical content of the document, then the processing unit's task is complete. On the other hand, if the topic belief vector does not adequately model the topical content of the document, then the processing unit determines how the topic belief vector should be modified to improve the prediction of modeling of the topical content.

  Recently Added Patents
Removable case
Methods and apparatus for dynamic identification (ID) assignment in wireless networks
Method and devices for handling access privileges
Timing controller capable of removing surge signal and display apparatus including the same
System and method for providing advice to consumer regarding a payment transaction
Identifying users of remote sessions
Method of targeting hydrophobic drugs to vascular lesions
  Randomly Featured Patents
Single piece pattern air bag
Mesh for a chair back rest
Preparation of bisphenols
Hybrid drive system
Dry shaver
Systems and methods to present web image search results for effective image browsing
Nonlinear blind demixing of single pixel underlying radiation sources and digital spectrum local thermometer
Method and apparatus for immunizing data in computer systems from corruption by assuming that incoming messages are corrupt unless proven valid
Structure and method for improving flow uniformity and reducing turbulence
Substituted aminoxyethyl sulfoxides and sulfones