IP Library Granted Patent US 8,620,944
Granted Patent B2
US 8,620,944 · App. 12/877,935 · Granted Dec 31, 2013

Systems and methods for keyword analyzer

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,620,944
App. No.
12/877,935
Granted
Dec 31, 2013
Kind
B2
Abstract

In one embodiment, a system and method is provided to browse and analyze files comprising text strings tagged with metadata. The system and method comprise various functions including browsing the metadata tags in the file, browsing the text strings, selecting subsets of the text strings by including or excluding strings tagged with specific metadata tags, selecting text strings by matching patterns of words and/or parts of speech in the text string and matching selected text strings to a database to identify similar text string. The system and method further provide functions to generate suggested text selection rules by analyzing a selected subset of a plurality of text strings.

Claims (72)

1. A method, comprising:

receiving a plurality of text strings, wherein each of the plurality of text strings has at least one of a plurality of metadata tags;

determining, via a data processing system, a count of a total number of text strings associated with each of the plurality of metadata tags;

displaying, via a user interface on a display device, each of the plurality of metadata tags and the respective count for the respective metadata tag;

receiving, via the user interface, a selected tag of the plurality of metadata tags;

identifying a string subset of the plurality of text strings associated with the selected tag;

identifying a tag subset of the plurality of metadata tags comprising all metadata tags associated with the string subset;

determining a revised count of a total number of text strings within the string subset associated with each of the tag subset;

displaying, via the user interface, each member of the tag subset along with its respective revised count;

receiving, via the user interface, an indication that the user wishes to display the string subset;

displaying, via the user interface, each of the text strings within the string subset;

receiving, via the user interface, an indication that the user wishes the interface to generate suggested rules for selecting text strings that match a pattern;

generating a plurality of suggested rules, wherein the plurality of suggested rules are generated by analyzing text and parts of speech patterns in the string subset, wherein each of the suggested rules matches at least one of the text strings within the string subset; and

displaying, via the user interface, each of the plurality of suggested rules.

2. The method of claim 1 , wherein the selected tag further comprises an indication that text strings associated with the metadata tag are to be excluded from the subset of the plurality of text strings, whereby each of the text strings within the string subset is not associated with the selected one of the plurality of metadata tags.

3. The method of claim 1 , further comprising:

receiving, via the user interface, a rule for selecting text strings that match a pattern;

identifying a string subset of the plurality of text strings, wherein each of the text strings within the string subset matches the rule for selecting text strings;

identifying a tag subset of the plurality of metadata tags comprising all metadata tags associated with the string subset; and

determining a revised count of a total number of text strings within the subset of the plurality of text strings associated with each of the subset of the plurality of metadata tags; and

displaying, via the user interface, each of the subset of the plurality of metadata tags and the respective revised count for the respective metadata tag.

4. The method of claim 1 , wherein the display of each of the text strings within the string subset comprises the respective text string, a last word of the respective text string and a representation of the parts of speech of the words comprising the respective text string.

5. The method of claim 4 , wherein the display of each of the text strings within the string subset further comprises user interface elements that allow the respective string to be flagged as desirable or undesirable.

6. The method of claim 1 , wherein each of the plurality of suggested rules is displayed in conjunction with a count of the number of text strings the respective rule matches in the subset of the plurality of text strings.

7. The method of claim 1 , wherein at least one of the plurality of suggested rules is a text pattern matching rule.

8. The method of claim 1 , wherein at least one of the plurality of suggested rules is a parts-of-speech pattern matching rule.

9. The method of claim 1 , further comprising:

receiving, via the user interface, an indication that at least a first string of the string subset represents desirable content;

receiving, via the user interface, an indication that the user wishes the interface to generate suggested rules for selecting text strings that match a pattern;

generating a plurality of suggested rules, wherein plurality of suggested rules are generated by analyzing text and parts of speech patterns in the at least a first string of the string subset, wherein each of the plurality of suggested rules matches the at least a first string of the string subset; and

displaying, via the user interface, each of the plurality of suggested rules.

10. The method of claim 9 , further comprising:

receiving, via the user interface, an indication that at least a second string of the string subset represents undesirable content,

wherein the plurality of suggested rules are generated by analyzing text and parts of speech patterns in the at least a first string of the string subset and the at least a second string of the string subset, and wherein each of the plurality of suggested rules matches the at least a first string of the string subset and does not match the at least a second string of the string subset.

11. The method of claim 1 , further comprising:

receiving, via the user interface, a selection of one of the text strings within the string subset;

receiving, via the user interface, an indication that the user wishes the interface to select similar text strings to the one of the text strings within the string subset; and

identifying, in a database of text strings, a plurality of similar text strings, wherein the plurality of similar text strings are identified by performing a full text search against the database of text strings using the one of the text strings within the string subset;

displaying, via a user interface on a display device, each of the plurality of similar text strings.

12. The method of claim 11 , further wherein the display of each of the plurality of similar text strings comprises the respective text string, a last word of the respective text string and a representation of the parts of speech of the words comprising the respective text string.

13. The method of claim 12 , wherein the display of each of the subset of the plurality of similar text strings further comprises user interface elements that allow the respective string to be flagged as desirable or undesirable.

14. A non-transitory machine-readable storage media embodying instructions, the instructions causing a data processing system to perform a method, the method comprising:

receiving a plurality of text strings, wherein each of the plurality of text strings has at least one of a plurality of metadata tags;

determining, via a data processing system, a count of a total number of text strings associated with each of the plurality of metadata tags;

displaying, via a user interface on a display device, each of the plurality of metadata tags and the respective count for the respective metadata tag;

receiving, via the user interface, a selection of one of the plurality of metadata tags;

receiving, via the user interface, a selected tag of the plurality of metadata tags;

identifying a string subset of the plurality of text strings associated with the selected tag;

identifying a tag subset of the plurality of metadata tags comprising all metadata tags associated with the string subset;

determining a revised count of a total number of text strings within the string subset associated with each of the tag subset;

displaying, via the user interface, each member of the tag subset along with its respective revised count;

receiving, via the user interface, an indication that the user wishes to display the string subset;

displaying, via the user interface, each of the text strings within the string subset;

receiving, via the user interface, an indication that the user wishes the interface to generate suggested rules for selecting text strings that match a pattern;

generating a plurality of suggested rules, wherein the plurality of suggested rules are generated by analyzing text and parts of speech patterns in the string subset, wherein each of the suggested rules matches at least one of the text strings within the string subset; and

displaying, via the user interface, each of the plurality of suggested rules.

15. A data processing system, comprising:

a display device to display a user interface;

a memory to store a plurality of text strings, wherein each of the plurality of text string has at least one of a plurality of metadata tags; and

a processor configured to:

receive a plurality of text strings, wherein each of the plurality of text strings has at least one of a plurality of metadata tags;

determine a count of a total number of text strings associated with each of the plurality of metadata tags;

display, via a user interface on the display device, each of the plurality of metadata tags and the respective count for the respective metadata tag;

receive, via the user interface, a selection of one of the plurality of metadata tags;

identify a string subset of the plurality of text strings associated with the selected tag;

identify a tag subset of the plurality of metadata tags comprising all metadata tags associated with the string subset;

determine a revised count of a total number of text strings within the string subset associated with each of the tag subset;

receive, via the user interface, an indication that the user wishes to display the string subset;

display, via the user interface, each of the text strings within the string subset;

receive, via the user interface, an indication that the user wishes the interface to generate suggested rules for selecting text strings that match a pattern;

generate a plurality of suggested rules, wherein the plurality of suggested rules are generated by analyzing text and parts of speech patterns in the string subset, wherein each of the suggested rules matches at least one of the text strings within the string subset; and

display, via the user interface, each of the plurality of suggested rules.

Assignments (7)
SECURITY INTEREST Recorded Nov 8, 2019
From: LEAF GROUP LTD.
To: SILICON VALLEY BANK
Reel/Frame 050957/0469 →
RELEASE OF SECURITY INTEREST Recorded Dec 13, 2016
From: OBSIDIAN AGENCY SERVICES, INC., AS AGENT
To: RIGHTSIDE OPERATING CO.
Reel/Frame 040725/0675 →
CHANGE OF NAME Recorded Nov 22, 2016
From: DEMAND MEDIA, INC.
To: LEAF GROUP LTD.
Reel/Frame 040730/0579 →
RELEASE OF INTELLECTUAL PROPERTY SECURITY INTEREST AT REEL/FRAME NO. 31123/0671 Recorded Nov 28, 2014
From: SILICON VALLEY BANK
To: DEMAND MEDIA, INC.
Reel/Frame 034494/0634 →
SECURITY INTEREST Recorded Aug 7, 2014
From: RIGHTSIDE OPERATING CO.
To: OBSIDIAN AGENCY SERVICES, INC.
Reel/Frame 033498/0848 →
RELEASE OF 2011 AND 2012 PATENT SECURITY INTERESTS Recorded Aug 29, 2013
From: SILICON VALLEY BANK, AS ADMINISTRATIVE AGENT
To: DEMAND MEDIA, INC.
Reel/Frame 031123/0458 →
SECURITY AGREEMENT Recorded Aug 29, 2013
From: DEMAND MEDIA, INC.
To: SILICON VALLEY BANK, AS ADMINISTRATIVE AGENT
Reel/Frame 031123/0671 →