Method and Apparatus For Classifying Digital Content Based on Ideological Bias of Authors
A method and apparatus for classifying a collection of digital documents based on ideological bias of authors. At least a portion of text of a digital document is received and parsed. Pairs of specific features text having specified relationships are detected. The pairs are then mapped to an ideological bias, based on an ideological bias ontology for example. Various actions can be taken on the digital documents based on the determined ideological bias.
1 . A method for classifying a collection of digital documents based on ideological bias of authors, the method comprising:
receiving at least a portion of text of a digital document;
parsing the portion of digital text;
detecting at least one pair of specific features of the portion of digital text having specified relationships;
mapping the at least pairs of specific features to an ideological bias based on the ideological bias ontology; and
taking action on the digital document based on the ideological bias.
2 . The method of claim 1 , wherein the relationships are specified by an ontology.
3 . The method of claim 1 , wherein said mapping step comprises scoring the at least pairs with a value relating to a specified ideological bias.
4 . The method of claim 2 , wherein the ontology includes entities and relations and the detecting step comprises detecting at least one entity and at least one relation as the at least one pair of specific features of the portion of the digital text having specified relationships.
5 . The method of claim 4 , wherein the ontology includes themes, each theme having at least one entity relation pairing.
6 . A computer architecture for classifying a collection of digital documents based on ideological bias of authors, the architecture comprising:
at least one processor; and
at least one memory operatively coupled to the at least one processor and storing instructions which, when executed by the processor, cause the processor to carry out the method of:
receiving at least a portion of text of a digital document;
parsing the portion of digital text;
detecting at least one pair of specific features of the portion of digital text having specified relationships;
mapping the at least pairs of specific features to an ideological bias based on the ideological bias ontology; and
taking action on the digital document based on the ideological bias.
7 . The architecture of claim 6 , wherein the relationships are specified by an ontology.
8 . The architecture of claim 6 , wherein said mapping step comprises scoring the at least pairs with a value relating to a specified ideological bias.
9 . The architecture of claim 7 , wherein the ontology includes entities and relations and the detecting step comprises detecting at least one entity and at least one relation as the at least one pair of specific features of the portion of the digital text having specified relationships.
10 . The architecture of claim 9 , wherein the ontology includes themes, each theme having at least one entity relation pairing.