Disclosed are a computer-readable code, system and method for classifying
a target document in the form of a digitally encoded natural-language text as belonging
to one or more of two or more different classes. Each of a plurality of non-generic
words and optionally, words groups characterizing the target document is selected
as a descriptive term if the term has an above-threshold selectivity value in at
least one library of texts in a field, where the selectivity value of a term is
a measure of the field-specificity of that term. There is then determined, for
each of the plurality of sample texts having associated classification identifiers,
a match score related to the number of descriptive terms present in or derived
from that text that match those in the target text. From the selected matched texts,
and the associated classification identifiers, a classification determination of
the target document is made.