IP Library Granted Patent US 8,818,788
Granted Patent B1
US 8,818,788 · App. 13/363,978 · Granted Aug 26, 2014

System, method and computer program product for identifying words within collection of text applicable to specific sentiment

Inventors: Dustin Mihalik (Cedar Park, TX); Dustin Friesenhahn (Austin, TX); Luveen Rupchand Wadhwani (Austin, TX)
Assignee: Bazaarvoice, Inc.
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,818,788
App. No.
13/363,978
Granted
Aug 26, 2014
Kind
B1
Abstract

A content intelligence module may implement a sentiment analysis method to identify words or phrases from user-generated content that are associated with a particular sentiment. The method may comprise grouping or splitting text into different sentiment segments, tokenizing words or phrases and/or removing stopwords across the sentiment segments, performing a frequency analysis to count the words or phrases in each sentiment segment, scaling the frequency results across the sentiment segments where necessary, and removing commonly used words from the sentiment segments. The words or phrases that are left in a specific sentiment segment are the most-used words for that sentiment segment. The word cloud module therefore allows for very quick generation of a summary around sentiment segments. A sentiment overview containing the summary can be presented to a user in connection with a selected product or service with which the user-generated content is associated.

Claims (44)

1. A method for analyzing sentiment, comprising:

at a first computer:

dividing a collection of text into a plurality of sentiment segments;

tokenizing words or phrases in the plurality of sentiment segments;

performing a frequency analysis on tokenized words or phrases in each sentiment segment of the plurality of sentiment segments;

performing a scaling operation to size individual sentiment segments based on results from the frequency analysis;

for each tokenized word or phrase in each sentiment segment of the plurality of sentiment segments, subtracting a first number of the tokenized word or phrase in the sentiment segment from a second number of the tokenized word or phrase in at least one other sentiment segment of the plurality of sentiment segments, thereby producing, for each sentiment segment of the plurality of sentiment segments, a list of words or phrases that apply specifically to the sentiment segment; and

providing the list of words or phrases that apply specifically to the sentiment segment to a second computer over a network connection.

2. The method of claim 1 , further comprising:

removing stopwords in each sentiment segment of the plurality of sentiment segments prior to performing the frequency analysis.

3. The method of claim 1 , wherein the collection of text is divided into the plurality of sentiment segments based on structured information about the text.

4. The method of claim 1 , wherein the collection of text comprises user-generated content.

5. The method of claim 1 , wherein the collection of text is associated with a product or service.

6. The method of claim 1 , wherein the scaling operation is performed based on an overall count of the tokenized words or phrases in each sentiment segment of the plurality of sentiment segments.

7. The method of claim 1 , wherein the plurality of sentiment segments comprises a positive sentiment and a negative sentiment.

8. A computer program product comprising at least one non-transitory computer readable medium storing instructions translatable by a first computer to:

divide a collection of text into a plurality of sentiment segments;

tokenize words or phrases in the plurality of sentiment segments;

perform a frequency analysis on tokenized words or phrases in each sentiment segment of the plurality of sentiment segments;

perform a scaling operation to size individual sentiment segments based on results from the frequency analysis;

for each tokenized word or phrase in each sentiment segment of the plurality of sentiment segments, subtract a first number of the tokenized word or phrase in the sentiment segment from a second number of the tokenized word or phrase in at least one other sentiment segment of the plurality of sentiment segments, thereby producing, for each sentiment segment of the plurality of sentiment segments, a list of words or phrases that apply specifically to the sentiment segment; and

provide the list of words or phrases that apply specifically to the sentiment segment to a second computer over a network connection.

9. The computer program product of claim 8 , wherein the instructions are further translatable by the first computer to perform:

removing stopwords in each sentiment segment of the plurality of sentiment segments prior to performing the frequency analysis.

10. The computer program product of claim 8 , wherein the collection of text is divided into the plurality of sentiment segments based on structured information about the text.

11. The computer program product of claim 8 , wherein the collection of text comprises user-generated content.

12. The computer program product of claim 8 , wherein the collection of text is associated with a product or service.

13. The computer program product of claim 8 , wherein the scaling operation is performed based on an overall count of the tokenized words or phrases in each sentiment segment of the plurality of sentiment segments.

14. A system, comprising:

at least one processor;

at least one non-transitory computer readable medium storing instructions translatable by the at least one processor to implement a word cloud module, the word cloud module being configured to:

divide a collection of text into a plurality of sentiment segments;

tokenize words or phrases in the plurality of sentiment segments;

perform a frequency analysis on tokenized words or phrases in each sentiment segment of the plurality of sentiment segments;

perform a scaling operation to size individual sentiment segments based on results from the frequency analysis;

for each tokenized word or phrase in each sentiment segment of the plurality of sentiment segments, subtract a first number of the tokenized word or phrase in the sentiment segment from a second number of the tokenized word or phrase in at least one other sentiment segment of the plurality of sentiment segments, thereby producing, for each sentiment segment of the plurality of sentiment segments, a list of words or phrases that apply specifically to the sentiment segment; and

provide the list of words or phrases that apply specifically to the sentiment segment to a second computer over a network connection.

15. The system of claim 14 , wherein the word cloud module is further configured to perform:

removing stopwords in each sentiment segment of the plurality of sentiment segments prior to performing the frequency analysis.

16. The system of claim 14 , wherein the collection of text is divided into the plurality of sentiment segments based on structured information about the text.

17. The system of claim 14 , wherein the collection of text comprises user-generated content.

18. The system of claim 14 , wherein the collection of text is associated with a product or service.

19. The system of claim 14 , wherein the scaling operation is performed based on an overall count of the tokenized words or phrases in each sentiment segment of the plurality of sentiment segments.

20. The system of claim 14 , wherein the plurality of sentiment segments comprises a positive sentiment and a negative sentiment.

Assignments (6)
RELEASE OF SECURITY INTEREST RECORDED AT 044804/0452 Recorded May 12, 2021
From: GOLUB CAPITAL MARKETS LLC
To: BAZAARVOICE, INC.
Reel/Frame 056208/0562 →
SECURITY INTEREST Recorded May 8, 2021
From: BAZAARVOICE, INC.; CURALATE, INC.
To: ALTER DOMUS (US) LLC
Reel/Frame 056179/0176 →
SECURITY INTEREST Recorded Feb 1, 2018
From: BAZAARVOICE, INC.
To: GOLUB CAPITAL MARKETS LLC
Reel/Frame 044804/0452 →
RELEASE OF SECURITY INTEREST Recorded Jan 18, 2018
From: COMERICA BANK
To: BAZAARVOICE, INC.
Reel/Frame 044659/0465 →
AMENDED AND RESTATED INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Nov 26, 2014
From: BAZAARVOICE, INC.
To: COMERICA BANK, AS AGENT
Reel/Frame 034474/0627 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 14, 2012
From: MIHALIK, DUSTIN; FRIESENHAHN, DUSTIN; WADHWANI, LUVEEN RUPCHAND
To: BAZAARVOICE, INC.
Reel/Frame 028199/0354 →