IP Library › Granted Patent US 11,238,234
Granted Patent B2
US 11,238,234 · App. 16/568,067 · Granted Feb 1, 2022

Adjusting a verbosity of a conversation turn

Inventors: Shun Jiang (San Jose, CA); Robert John Moore (San Jose, CA); Margaret Helen Szymanski (Santa Clara, CA); Lei Huang (Mountain View, CA); Guangjie Ren (Belmont, CA); Peifeng Yin (San Jose, CA)
Assignee: International Business Machines Corporation
G06F40/30G06F16/3329G06N20/00G06N3/084
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,238,234
App. No.
16/568,067
Granted
Feb 1, 2022
Kind
B2
Abstract

In one general aspect, a computer-implemented method includes identifying current choices with different verbosity levels for a current turn in a conversation; normalizing multi-dimensional verbosity vectors for each of the current choices to obtain a normalized value for each of the current choices; determining a state definition for the current turn in the conversation, utilizing the normalized values for each of the current choices; providing the state definition for the current turn in the conversation and the normalized values for each of the current choices to a trained reinforcement learning module; receiving, from the trained reinforcement learning module, a score associated with each of the current choices for the current turn in the conversation; and selecting one of the current choices to be entered for the current turn in the conversation, based on the score associated with each of the current choices for the current turn in the conversation.

Claims (40)

1. A computer-implemented method, comprising:

identifying current choices with different verbosity levels for a current turn in a conversation;

normalizing multi-dimensional verbosity vectors for each of the current choices to obtain a normalized value for each of the current choices;

determining a state definition for the current turn in the conversation, utilizing the normalized values for each of the current choices;

providing the state definition for the current turn in the conversation and the normalized values for each of the current choices to a trained reinforcement learning module;

receiving, from the trained reinforcement learning module, a score associated with each of the current choices for the current turn in the conversation; and

selecting one of the current choices to be entered for the current turn in the conversation, based on the score associated with each of the current choices for the current turn in the conversation.

2. The computer-implemented method of claim 1 , wherein the conversation includes a question and answer (Q and A) conversation.

3. The computer-implemented method of claim 1 , wherein each of the current choices includes a textual word or phrase having wording and vocabulary different from all other current choices.

4. The computer-implemented method of claim 1 , wherein each of the current choices includes a multi-dimensional verbosity vector describing a verbosity of the choice.

5. The computer-implemented method of claim 1 , wherein the multi-dimensional verbosity vectors for each current choice each include a value for a number of words used within the current choice.

6. The computer-implemented method of claim 1 , wherein the multi-dimensional verbosity vectors for each current choice each include a value for a type of words used within the current choice.

7. The computer-implemented method of claim 1 , wherein normalizing the multi-dimensional verbosity vectors for each of the current choices includes normalizing a word count for each of the current choices.

8. The computer-implemented method of claim 1 , wherein normalizing the multi-dimensional verbosity vectors for each of the current choices includes normalizing one or more domain terminologies for each of the current choices.

9. The computer-implemented method of claim 1 , wherein the state definition is determined by summing an absolute distance of each of a plurality of verbosity vectors of past turns in the conversation.

10. The computer-implemented method of claim 1 , wherein the state definition is determined by calculating a difference between a normalized verbosity value for a last turn in the conversation and the normalized value for each of the current choices.

11. The computer-implemented method of claim 1 , wherein the trained reinforcement learning module includes a Q-learning method.

12. The computer-implemented method of claim 1 , wherein for each of the current choices, the score includes a value indicating a suitability of the current choice for use within the current turn in the conversation.

13. The computer-implemented method of claim 1 , wherein a current choice having a highest score is determined to be least likely to confuse a user in the conversation, while minimizing a verbosity level, and is selected to be entered for the current turn in the conversation.

14. The computer-implemented method of claim 1 , wherein the state definition is determined by summing an absolute distance of each of a plurality of verbosity vectors.

15. A computer program product for adjusting a verbosity of a conversation turn, the computer program product comprising a computer readable storage medium having program instructions embodied therewith, wherein the computer readable storage medium is not a transitory signal per se, the program instructions executable by a processor to cause the processor to perform a method comprising:

identifying, by the processor, current choices with different verbosity levels for a current turn in a conversation;

normalizing, by the processor, multi-dimensional verbosity vectors for each of the current choices to obtain a normalized value for each of the current choices;

determining, by the processor, a state definition for the current turn in the conversation, utilizing the normalized values for each of the current choices;

providing, by the processor, the state definition for the current turn in the conversation and the normalized values for each of the current choices to a trained reinforcement learning module;

receiving, from the trained reinforcement learning module by the processor, a score associated with each of the current choices for the current turn in the conversation; and

selecting, by the processor, one of the current choices to be entered for the current turn in the conversation, based on the score associated with each of the current choices for the current turn in the conversation.

16. The computer program product of claim 15 , wherein the conversation includes a question and answer (Q and A) conversation.

17. The computer program product of claim 15 , wherein each of the current choices includes a textual word or phrase having wording and vocabulary different from the other current choices.

18. The computer program product of claim 15 , wherein each of the current choices includes a multi-dimensional verbosity vector describing a verbosity of the choice.

19. The computer program product of claim 15 , wherein the multi-dimensional verbosity vectors for each current choice each include a value for a number of words used within the current choice.

20. A system, comprising:

a processor; and

logic integrated with the processor, executable by the processor, or integrated with and executable by the processor, the logic being configured to:

identify current choices with different verbosity levels for a current turn in a conversation;

normalize multi-dimensional verbosity vectors for each of the current choices to obtain a normalized value for each of the current choices;

determine a state definition for the current turn in the conversation, utilizing the normalized values for each of the current choices;

provide the state definition for the current turn in the conversation and the normalized values for each of the current choices to a trained reinforcement learning module;

receive, from the trained reinforcement learning module, a score associated with each of the current choices for the current turn in the conversation; and

select one of the current choices to be entered for the current turn in the conversation, based on the score associated with each of the current choices for the current turn in the conversation.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 4, 2019
From: JIANG, SHUN; MOORE, ROBERT JOHN; SZYMANSKI, MARGARET HELEN; HUANG, LEI; REN, GUANGJIE; YIN, PEIFENG
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 050626/0778 →
Continuity (1)
Related Publication 20210073337A1 · Mar 11, 2021
Cited By (1)
US 12,292,905