Text-classification system and method

Disclosed are a computer-readable code, system and method for classifying a target document in the form of a digitally encoded natural-language text as belonging to one or more of two or more different classes. Each of a plurality of non-generic words and optionally, words groups characterizing the target document is selected as a descriptive term if the term has an above-threshold selectivity value in at least one library of texts in a field, where the selectivity value of a term is a measure of the field-specificity of that term. There is then determined, for each of the plurality of sample texts having associated classification identifiers, a match score related to the number of descriptive terms present in or derived from that text that match those in the target text. From the selected matched texts, and the associated classification identifiers, a classification determination of the target document is made.

Web www.patentalert.com

< Methods for establishing a pathways database and performing pathway searches

< Methods for accumulating selenium in edible brassica

> Modified Cry3A toxins and nucleic acid sequences coding therefor

> System and methods for noninvasive electrocardiographic imaging (ECGI) using generalized minimum residual (GMRes)

~ 00247