Nunberg, Geoffrey D., Jan O. Pedersen, Hinrich Schütze, Brett Kessler & Gregory Grefenstette. 1999. Text genre identification. European Patent EP 0 889 417. London, England: European Patent Office.

Abstract

A processor implemented method of identifying the genre of a machine readable, untagged text. The processor implemented method begins by generating a cue vector from the text, which represents occurrences in the text of a first set of nonstructural, surface cues, which are easily computable. Afterward, the processor determines whether the text is an instance of a first text genre using the cue vector and a weighting vector associated with the first text genre.

Application

APA citation:

Nunberg, G. D., Pedersen, J. O., Schütze, H., Kessler, B., & Grefenstette, G. (1999). European Patent EP 0 889 417. London, England: European Patent Office.


Last change 2009-08-07T00:41:18-0500