IP Library › Granted Patent US 10,650,034
Granted Patent B2
US 10,650,034 · App. 16/102,875 · Granted May 12, 2020

Categorizing users based on similarity of posed questions, answers and supporting evidence

Inventors: Christopher S. Alkov (Austin, TX); Suzanne L. Estrada (Boca Raton, FL); Peter F. Haggar (Raleigh, NC); Kevin B. Haverlock (Cary, NC)
Assignee: International Business Machines Corporation
G06F16/353G06F16/243G06F16/2465G06F16/285G06F16/287G06F16/3329G06N20/00G06Q30/0251G06Q50/00H04L51/00H04L51/04H04L51/046H04L51/32G06Q50/265
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,650,034
App. No.
16/102,875
Granted
May 12, 2020
Kind
B2
Abstract

Mechanisms are provided for performing an operation based on an identification of similar lines of questioning by input question sources. Question information identifying extracted features of an input question and a first source of the input question is obtained. A clustering operation is performed to cluster the input question with one or more other questions of a cluster based on a similarity of the extracted features of the input question to features of the one or more other questions. An operation is performed based on results of the clustering of the input question with the one or more other questions.

Claims (37)

1. A method, in a data processing system comprising a processor and a memory, for performing an operation based on an identification of similar lines of questioning by input question sources, the method comprising:

obtaining, by the data processing system, question information identifying extracted features of an input question and a first source of the input question;

performing, by the data processing system, a clustering operation to cluster the input question with one or more other questions of a cluster based on a similarity of the extracted features of the input question to features of the one or more other questions; and

performing, by the data processing system, an operation based on results of the clustering of the input question with the one or more other questions, wherein the operation comprises at least one of automatically initiating a collaboration between the first source of the input question and a second source of another question in the cluster via at least one computer implemented collaboration system, or automatically initiating a communication between the first source and the second source at least by automatically establishing a communication connection between the first source and the second source, wherein the input question is a question input to a Question and Answer (QA) system which processes the input question to generate an answer to the input question based on a corpus of information, and further generates one or more supporting evidence passages supporting the answer as being a correct answer for the input question, and wherein performing the clustering operation to cluster the input question with the one or more other questions of a cluster comprises performing the clustering based on features of the input question, features of an answer, and features of the one or more supporting evidence passages.

2. The method of claim 1 , wherein, as part of the clustering operation, a plurality of different weighting factors are established for different features of at least one of the input question, the answer, or the one or more supporting evidence passages for use in performing the clustering operation.

3. The method of claim 1 , wherein:

the first source and second source are communication devices associated with users,

the operation comprises sending a request to at least one of the first source or the second source requesting that the corresponding first source or second source join in a collaborative communication with the other source via the collaborative communication system, and

the collaborative communication is at least one of a chat group, an instant messaging session, an electronic mail exchange, or a conference telephone call.

4. The method of claim 1 , wherein the first source and second source are communication devices associated with users, and wherein the operation further comprises at least one of:

sending targeted advertising, by a third party source, to the first source and the second source,

providing information about the first source and the second source to the third party source,

initiating a targeted processing of user information of the first source or second source to analyze other activities by the first source or second source, and

identifying, by a government organization system, the first source and the second source as potentially engaged in illegal activity and targeting the first source and second source for further investigation of illegal activity.

5. The method of claim 1 , wherein the extracted features and features of the one or more other questions comprise at least one of a focus, lexical answer type (LAT), question classification (QClass), or question sections (QSection).

6. The method of claim 1 , wherein performing the clustering operation comprising clustering the input question based on a plot of the extracted features of the input question relative to a plot of features of the one or more other questions using at least one threshold value specifying a required distance from the plot of the extracted features of the input question to the plot of the features of the one or more other questions in order for the input question and the one or more other questions to be considered part of a same cluster.

7. A computer program product comprising a computer readable storage medium having a computer readable program stored therein, wherein the computer readable program, when executed on a computing device, causes the computing device to:

obtain question information identifying extracted features of an input question and a first source of the input question;

perform a clustering operation to cluster the input question with one or more other questions of a cluster based on a similarity of the extracted features of the input question to features of the one or more other questions; and

perform an operation based on results of the clustering of the input question with the one or more other questions, wherein the operation comprises at least one of automatically initiating a collaboration between the first source of the input question and a second source of another question in the cluster via at least one computer implemented collaboration system, or automatically initiating a communication between the first source and the second source at least by automatically establishing a communication connection between the first source and the second source, wherein the input question is a question input to a Question and Answer (QA) system which processes the input question to generate an answer to the input question based on a corpus of information, and further generates one or more supporting evidence passages supporting the answer as being a correct answer for the input question, and wherein the computer readable program further causes the computing device to perform the clustering operation to cluster the input question with the one or more other questions of a cluster at least by performing the clustering based on features of the input question, features of an answer, and features of the one or more supporting evidence passages.

8. The computer program product of claim 7 , wherein, as part of the clustering operation, a plurality of different weighting factors are established for different features of at least one of the input question, the answer, or the one or more supporting evidence passages for use in performing the clustering operation.

9. The computer program product of claim 7 , wherein:

the first source and second source are communication devices associated with users,

the operation comprises sending a request to at least one of the first source or the second source requesting that the corresponding first source or second source join in a collaborative communication with the other source via the collaborative communication system, and

the collaborative communication is at least one of a chat group, an instant messaging session, an electronic mail exchange, or a conference telephone call.

10. The computer program product of claim 7 , wherein the first source and second source are communication devices associated with users, and wherein the operation further comprises at least one of:

sending targeted advertising, by a third party source, to the first source and the second source,

providing information about the first source and the second source to the third party source,

initiating a targeted processing of user information of the first source or second source to analyze other activities by the first source or second source, and

identifying, by a government organization system, the first source and the second source as potentially engaged in illegal activity and targeting the first source and second source for further investigation of illegal activity.

11. The computer program product of claim 7 , wherein the extracted features and features of the one or more other questions comprise at least one of a focus, lexical answer type (LAT), question classification (QClass), or question sections (QSection).

12. An apparatus comprising:

a processor; and

a memory coupled to the processor, wherein the memory comprises instructions which, when executed by the processor, cause the processor to:

obtain question information identifying extracted features of an input question and a first source of the input question;

perform a clustering operation to cluster the input question with one or more other questions of a cluster based on a similarity of the extracted features of the input question to features of the one or more other questions; and

perform an operation based on results of the clustering of the input question with the one or more other questions, wherein the operation comprises at least one of automatically initiating a collaboration between the first source of the input question and a second source of another question in the cluster via at least one computer implemented collaboration system, or automatically initiating a communication between the first source and the second source at least by automatically establishing a communication connection between the first source and the second source, wherein the input question is a question input to a Question and Answer (QA) system which processes the input question to generate an answer to the input question based on a corpus of information, and further generates one or more supporting evidence passages supporting the answer as being a correct answer for the input question, and wherein the instructions further cause the processor to perform the clustering operation to cluster the input question with the one or more other questions of a cluster at least by performing the clustering based on features of the input question, features of an answer, and features of the one or more supporting evidence passages.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 14, 2018
From: ALKOV, CHRISTOPHER S.; ESTRADA, SUZANNE L.; HAGGAR, PETER F.; HAVERLOCK, KEVIN B.
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 046787/0572 →
Continuity (3)
Continuation 15405500 · Jan 13, 2017
Continuation 14267184 · May 1, 2014
Related Publication 20190005127A1 · Jan 3, 2019