Systems and methods for calculating category proportions
Abstract
Systems and methods are provided for classifying text based on language using one or more computer servers and storage devices. A computer-implemented method includes receiving a training set of elements, each element in the training set being assigned to one of a plurality of categories and having one of a plurality of content profiles associated therewith; receiving a population set of elements, each element in the population set having one of the plurality of content profiles associated therewith; and calculating using at least one of a stacked regression algorithm, a bias formula algorithm, a noise elimination algorithm, and an ensemble method consisting of a plurality of algorithmic methods the results of which are averaged, based on the content profiles associated with and the categories assigned to elements in the training set and the content profiles associated with the elements of the population set, a distribution of elements of the population set over the categories.
Claims
exact text as granted — not AI-modifiedWhat is claimed is:
1 . A computer-implemented method performed by a computer processor, comprising:
(a) receiving by the computer processor a training set of elements, each element in the training set being assigned to one of a plurality of categories and having one of a plurality of content profiles associated therewith; (b) receiving by the computer processor a population set of elements, each element in the population set having one of the plurality of content profiles associated therewith; and (c) calculating by the computer processor applying the stacked regression method, based on the content profiles associated with and the categories assigned to elements in the training set and the content profiles associated with the elements of the population set, a distribution of elements of the population set over the categories.
2 . A computer-implemented method performed by a computer processor, comprising:
(a) receiving by the computer processor a training set of elements, each element in the training set being assigned to one of a plurality of categories and having one of a plurality of content profiles associated therewith; (b) receiving by the computer processor a population set of elements, each element in the population set having one of the plurality of content profiles associated therewith; and (c) calculating by the computer processor applying the bias formula method, based on the content profiles associated with and the categories assigned to elements in the training set and the content profiles associated with the elements of the population set, a distribution of elements of the population set over the categories.
3 . A computer-implemented method performed by a computer processor, comprising:
(a) receiving by the computer processor a training set of elements, each element in the training set being assigned to one of a plurality of categories and having one of a plurality of content profiles associated therewith; (b) receiving by the computer processor a population set of elements, each element in the population set having one of the plurality of content profiles associated therewith; and (c) calculating by the computer processor applying the noise elimination method, based on the content profiles associated with and the categories assigned to elements in the training set and the content profiles associated with the elements of the population set, a distribution of elements of the population set over the categories.
4 - 12 . (canceled)Join the waitlist — get patent alerts
Track US2017046630A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.