IP Library Granted Patent US 10,331,714
Granted Patent B2
US 10,331,714 · App. 14/079,406 · Granted Jun 25, 2019

Methods, systems, and computer-readable media for semantically enriching content and for semantic navigation

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,331,714
App. No.
14/079,406
Granted
Jun 25, 2019
Kind
B2
Abstract

Content of different formats may be sourced from various data sources such as content servers and ingested into a data integration server by an ingestion broker embodied on a non-transitory computer readable medium. The ingestion broker may normalize the content of different formats into a uniform representation that can be indexed and delivered across multiple digital channels for a variety of applications. The normalized content may be analyzed and semantic metadata may be determined from the normalized content. The normalized content can be semantically enriched by associating the semantic metadata and the like with the content. The semantic metadata can be stored in a semantic index that can be used for searching via the data integration server. During search, the semantic metadata can be instantiated as facets for user navigation and refinement of search criteria and additional semantic relationships can be assigned to the words in the normalized content.

Claims (49)

1. A method, comprising:

performing, by a computing device having a processor and a non-transitory computer readable medium, an ingestion process to ingest content of different formats sourced from a plurality of data sources into search indices, the ingestion process including:

inputting into an ingestion broker embodied on the non-transitory computer readable medium, the content of different formats sourced from a plurality of data sources, the ingestion broker supporting chaining of processors such that, during the ingestion process and prior to any indexing process, an indexing engine can call a content analytics module or a semantic annotator that semantically analyzes, annotates, and relates data in the content before the content is indexed;

normalizing, by the ingestion broker, the content of different formats into a uniform representation that is supported by the indexing engine and that is deliverable across multiple digital channels for a variety of applications;

calling the content analytics module or the semantic annotator, the calling performed by the indexing engine; and

analyzing the content having the uniform representation, the analyzing performed by the content analytics module or the semantic annotator during the ingestion process and prior to any indexing process, the analyzing including:

extracting information from a syntactic structure reflecting semantics of the content having the uniform representation;

semantically mapping the information extracted from the syntactic structure with information about meanings of words in the content having the uniform representation in order to determine semantic metadata; and

associating the semantic metadata with the words in the content having the uniform representation; and

subsequent to the ingestion process, indexing the content having the uniform representation, the indexing performed by the indexing engine based on both syntax and the semantic metadata determined by the content analytics module or the semantic annotator during the ingestion process, the indexing including storing the semantic metadata in a semantic index, wherein the content having the uniform representation is searchable via the semantic index by association with the semantic metadata stored in the semantic index.

2. The method of claim 1 , wherein the analyzing further comprises classifying the content having the uniform representation.

3. The method of claim 1 , wherein the analyzing further comprises determining annotations from the content having the uniform representation and wherein the semantic metadata associated with the content having the uniform representation comprises the annotations.

4. The method of claim 1 , wherein the analyzing further comprises performing a sentiment analysis on subjectivity and tonality of the content having the uniform representation.

5. The method of claim 1 , wherein the ingestion process further comprises assigning, by the content analytics module or the semantic annotator, semantic relationships to words in the content having a uniform representation.

6. The method of claim 5 , wherein additional semantic relationships are assigned to the words during user search.

7. The method of claim 1 , wherein the semantic index is part of a search index.

8. A computer program product comprising a non-transitory computer readable medium storing instructions translatable by a data integration server to:

perform an ingestion process to ingest content of different formats sourced from a plurality of data sources into search indices, the data integration server supporting chaining of processors such that, during the ingestion process and prior to any indexing process, an indexing engine can call a content analytics module or a semantic annotator that semantically analyzes, annotates, and relates data in the content before the content is indexed, the ingestion process including:

receiving content of different formats sourced from a plurality of data sources;

normalizing the content of different formats into a uniform representation that is supported by the indexing engine of the data integration server and that is deliverable across multiple digital channels for a variety of applications;

calling the content analytics module or the semantic annotator; and

analyzing the content having the uniform representation, the analyzing performed by the content analytics module or the semantic annotator during the ingestion process and prior to any indexing process, the analyzing including:

extracting information from a syntactic structure reflecting semantics of the content having the uniform representation;

semantically mapping the information extracted from the syntactic structure with information about meanings of words in the content having the uniform representation in order to determine semantic metadata; and

associating the semantic metadata with the words in the content having the uniform representation; and

subsequent to performing the ingestion process, indexing the content having the uniform representation, the indexing performed based on both syntax and the semantic metadata determined during the ingestion process, the indexing including storing the semantic metadata in a semantic index, wherein the content having the uniform representation is searchable via the semantic index by association with the semantic metadata stored in the semantic index.

9. The computer program product of claim 8 , wherein the analyzing further comprises classifying the content having the uniform representation.

10. The computer program product of claim 8 , wherein the analyzing further comprises determining annotations from the content having the uniform representation and wherein the semantic metadata associated with the content having the uniform representation comprises the annotations.

11. The computer program product of claim 8 , wherein the analyzing further comprises performing a sentiment analysis on subjectivity and tonality of the content having the uniform representation.

12. The computer program product of claim 8 , wherein the ingestion process further comprises assigning, by the content analytics module or the semantic annotator, semantic relationships to words in the content having uniform representation.

13. The computer program product of claim 12 , wherein additional semantic relationships are assigned to the words during user search.

14. The computer program product of claim 8 , wherein the semantic index is part of a search index.

15. A system, comprising:

a processor; and

a non-transitory computer readable medium storing instructions translatable by the processor to:

perform an ingestion process to ingest content of different formats sourced from a plurality of data sources into search indices, the system supporting chaining of processors such that, during the ingestion process and prior to any indexing process, an indexing engine can call a content analytics module or a semantic annotator that semantically analyzes, annotates, and relates data in the content before the content is indexed, the ingestion process including:

receiving content of different formats sourced from a plurality of data sources;

normalizing the content of different formats into a uniform representation that is supported by the indexing engine and that is deliverable across multiple digital channels for a variety of applications;

calling the content analytics module or the semantic annotator; and

analyzing the content having the uniform representation, the analyzing performed by the content analytics module or the semantic annotator during the ingestion process and prior to any indexing process, the analyzing including:

extracting information from a syntactic structure reflecting semantics of the content having the uniform representation;

semantically mapping the information extracted from the syntactic structure with information about meanings of words in the content having the uniform representation in order to determine semantic metadata; and

associating the semantic metadata with the words in the content having the uniform representation; and

subsequent to performing the ingestion process, indexing the content having the uniform representation, the indexing performed based on both syntax and the semantic metadata determined during the ingestion process, the indexing including storing the semantic metadata in a semantic index, wherein the content having the uniform representation is searchable via the semantic index by association with the semantic metadata stored in the semantic index.

16. The system of claim 15 , wherein the analyzing further comprises classifying the content having the uniform representation.

17. The system of claim 15 , wherein the analyzing further comprises determining annotations from the content having the uniform representation and wherein the semantic metadata associated with the content having the uniform representation comprises the annotations.

18. The system of claim 15 , wherein the analyzing further comprises performing a sentiment analysis on subjectivity and tonality of the content having the uniform representation.

19. The system of claim 15 , wherein the ingestion process further comprises assigning, by the content analytics module or the semantic annotator, semantic relationships to words in the content having uniform representation.

20. The system of claim 19 , wherein additional semantic relationships are assigned to the words during user search.

Assignments (5)
IP BUSINESS SALE AGREEMENT Recorded Aug 30, 2016
From: OPEN TEXT S.A.
To: OT IP SUB, LLC
Reel/Frame 039872/0605 →
CERTIFICATE OF AMALGAMATION Recorded Aug 30, 2016
From: IP OT SUB ULC
To: OPEN TEXT SA ULC
Reel/Frame 039872/0662 →
CERTIFICATE OF CONTINUANCE Recorded Aug 30, 2016
From: OT IP SUB, LLC
To: IP OT SUB ULC
Reel/Frame 039986/0689 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 9, 2014
From: OPEN TEXT CORPORATION
To: OPEN TEXT S.A.
Reel/Frame 032863/0520 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 7, 2014
From: DIMASSIMO, PASCAL; PETTIGREW, STEVE; BROUSSEAU, MARTIN; SIMARD, CHARLES-OLIVIER; WILLIAMS, ERIC; LACROIX, FRANCIS; DOWGAILENKO, ALEX; DELIGIA, AGOSTINO; TEXIER, JEAN-MICHEL
To: OPEN TEXT CORPORATION
Reel/Frame 032839/0897 →