IP Library Granted Patent US 8,019,760
Granted Patent B2
US 8,019,760 · App. 11/774,908 · Granted Sep 13, 2011

Clustering system and method

Assignee: Vivisimo, Inc.
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,019,760
App. No.
11/774,908
Granted
Sep 13, 2011
Kind
B2
Abstract

An increase in information available to a user of computing technologies has a tendency to increase the number of topics that are similarly related. Given the large amount of information that is now available, it is increasingly likely that a first set of search results generated in response to an initial search query will contain information that is not of interest to the user. What is needed in the art is a technique to enable a search query to be conducted by taking advantage of linguistic feedback. Furthermore, what is needed is a technique to enable the presentation of search results to be refined in a manner based on what is not of interest to a user, either intrinsically or because the user has already seen and evaluated certain information and next wants to see more or different information.

Claims (45)

1. A computer implemented method, comprising:

storing search results at a server based on a search engine query, wherein said search results comprise a plurality of items;

generating at the server a first set of clusters responsive to the search engine query, wherein each of said items is associated with at least one cluster in said first set of clusters;

sending the first set of clusters to a terminal configured to display the first set of clusters;

receiving at the server user input consisting of an indication to recluster the search results;

generating at the server a second set of clusters, wherein said second set of clusters excludes one or more clusters from said first set of clusters, and wherein each of said items is associated with at least one cluster in said second set of clusters; and

sending the second set of clusters to the terminal configured to display the second set of clusters,

wherein each cluster is defined by a cluster title, and wherein generating at the server a second set of clusters comprises excluding from the second set of clusters one or more cluster titles used in said first set of clusters, excluding the literal phraseology of at least one cluster title in the first set of clusters from the second set of clusters, and excluding a linguistic equivalence class corresponding to at least one cluster title in the first set of clusters from the second set of clusters, and

wherein generating the second set of clusters comprises excluding each displayed cluster of the first set of clusters from the second set of clusters.

2. The computer implemented method of claim 1 , wherein each generating step comprises:

determining one or more linguistic equivalence classes, each linguistic equivalence class identifying a primary term and one or more corresponding linguistically similar terms, and

for each linguistic equivalence class, treating all linguistically similar terms within the search results as identical to the corresponding primary term.

3. The computer implemented method of claim 1 , wherein generating the second set of clusters comprises allowing the second set of clusters to use a portion of a title of at least one cluster within the first set of clusters.

4. The computer implemented method of claim 1 , wherein the steps of generating each of the first and second set of clusters excludes clusters that would otherwise only include stopwords.

5. An apparatus comprising:

a processor; and

memory storing computer executable instructions that, when executed by the processor, perform a method of clustering, comprising:

storing search results based on a search query, wherein said search results comprise a plurality of items;

generating a first set of clusters responsive to the search query, wherein each of said items is associated with at least one cluster in said first set of clusters;

sending the first set of clusters to a terminal configured to display the first set of clusters;

receiving user input consisting of an indication to recluster the search results;

generating a second set of clusters, wherein said second set of clusters excludes one or more clusters from said first set of clusters, and wherein each of said items is associated with at least one cluster in said second set of clusters; and

sending the second set of clusters to the terminal configured to display the second set of clusters,

wherein each cluster is defined by a cluster title, and wherein generating at the server a second set of clusters comprises excluding from the second set of clusters one or more cluster titles used in said first set of clusters, excluding the literal phraseology of at least one cluster title in the first set of clusters from the second set of clusters, and excluding a linguistic equivalence class corresponding to at least one cluster title in the first set of clusters from the second set of clusters, and

wherein generating the second set of clusters comprises excluding each displayed cluster of the first set of clusters from the second set of clusters.

6. The apparatus of claim 5 , wherein each generating step comprises:

determining one or more linguistic equivalence classes, each linguistic equivalence class identifying a primary term and one or more corresponding linguistically similar terms, and

for each linguistic equivalence class, treating all linguistically similar terms within the search results as identical to the corresponding primary term.

7. The apparatus of claim 5 , further including instructions enabling the search query to be refined.

8. A data processing system comprising:

a processor; and

memory storing computer executable instructions that, when executed by the processor, perform a clustering method comprising:

receiving search results, wherein said search results comprise a plurality of items;

receiving a first set of clusters generated by a server, wherein each of said items is associated with at least one cluster in said first set of clusters;

displaying the first set of clusters to a user;

sending to the server user input consisting of an indication to recluster the search results;

receiving a second set of clusters generated by the server, wherein said second set of clusters excludes one or more clusters from said first set of clusters, and wherein each of said items is associated with at least one cluster in said second set of clusters; and

displaying the second set of clusters to the user,

wherein each cluster is defined by a cluster title, and wherein generating at the server a second set of clusters comprises excluding from the second set of clusters one or more cluster titles used in said first set of clusters, excluding the literal phraseology of at least one cluster title in the first set of clusters from the second set of clusters, and excluding a linguistic equivalence class corresponding to at least one cluster title in the first set of clusters from the second set of clusters, and

wherein generating the second set of clusters comprises excluding each displayed cluster of the first set of clusters from the second set of clusters.

9. The system of claim 8 , wherein the server, in generating each of the first and second sets of clusters performs the steps of:

determining one or more linguistic equivalence classes, each linguistic equivalence class identifying a primary term and one or more corresponding linguistically similar terms, and

for each linguistic equivalence class, treating all linguistically similar terms within the search results as identical to the corresponding primary term.

10. The system of claim 8 , wherein the server, in generating the second set of clusters excludes a theme using a heuristic that infers the theme has been viewed by a user.

11. The system of claim 8 , wherein the search results are identified based on an informal search query.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 29, 2013
From: VIVISIMO, INC.
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 029709/0365 →
RELEASE Recorded May 10, 2012
From: SILICON VALLEY BANK
To: VIVISIMO, INC.
Reel/Frame 028197/0855 →
SECURITY AGREEMENT Recorded May 3, 2011
From: VIVISIMO, INC.
To: SILICON VALLEY BANK
Reel/Frame 026214/0471 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 9, 2007
From: VALDES-PEREZ, RAUL E.; DOS SANTOS LESSA, ANDRE; PALMER, CHRISTOPHER ROBERT; PESENTI, JEROME
To: VIVISIMO, INC.
Reel/Frame 019532/0134 →
Continuity (1)
Related Publication 20090019026A1 · Jan 15, 2009