IP Library › Granted Patent US 9,984,376
Granted Patent B2
US 9,984,376 · App. 15/087,555 · Granted May 29, 2018

Method and system for automatically identifying issues in one or more tickets of an organization

Inventors: Venkatakrishnan Rajaram (Bengaluru, IN); Narayanan Ramani Konnayar (Bengaluru, IN); Ria Chakraborty (Kolkata, IN); Malathi Bellam Soundararajan (Bangalore, IN)
Assignee: WIPRO LIMITED
G06Q30/016G06F17/271G06F17/2725G06F17/2775
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,984,376
App. No.
15/087,555
Granted
May 29, 2018
Kind
B2
Abstract

The present disclosure relates to method and system for automatically identifying one or more issues in one or more tickets of an organization. An issue identification system retrieves a sequence pattern from ticket data received from one or more data sources. The issue identification system generates one or more first sub-sequence patterns of the n-grams from the sequence pattern. Further, frequency of occurrence and Part-of-Speech (POS) weightage of each of the one or more first sub-sequence patterns of the n-grams are determined by the issue identification system. A first score is determined for each of the one or more first sub-sequence patterns of the n-grams based on both the frequency and the POS weightage. Upon determining the first score, the issue identification system identifies one or more issues in the one or more tickets automatically based on the first sub-sequence pattern of the n-grams associated with a highest first score.

Claims (63)

1. A method for automatically identifying one or more issues in one or more tickets of an organization, the method comprising:

receiving, by an issue identification system, ticket data of one or more tickets related to a service category from one or more data sources;

generating, by the issue identification system, one or more first sub-sequence patterns of n-grams for the one or more tickets from a sequence pattern retrieved from the ticket data;

determining, by the issue identification system, a frequency of occurrence of each of the one or more first sub-sequence patterns of the n-grams and a Part-of-Speech (POS) weightage of the one or more first sub-sequence patterns of the n-grams;

determining, by the issue identification system, a first score for each of the one or more first sub-sequence patterns of the n-grams based on the frequency of occurrence and the POS weightage;

identifying, by the issue identification system, automatically, one or more issues in the one or more tickets based on the first sub-sequence pattern of the n-grams and the first score;

identifying, by the issue identification system, the sequence pattern corresponding to the one or more first sub-sequence patterns of the n-grams having the first score less than a predefined value for each of the one or more tickets;

generating, by the issue identification system, one or more second sub-sequence patterns of the n-grams by removing one or more words in the sequence pattern in order of occurrence, wherein a distance value is associated with each of the one or more second sub-sequence patterns based on the one or more words removed in the sequence pattern;

determining, by the issue identification system, a frequency of occurrence of each of the one or more second sub-sequence patterns of the n-grams and a POS weightage of the one or more second sub-sequence patterns of the n-grams;

determining, by the issue identification system, a second score for each of the one or more second sub-sequence patterns of the n-grams based on the frequency of occurrence and the POS weightage of the one or more second sub-sequence patterns of the n-grams; and

updating, by the issue identification system, automatically, the one or more issues in the one or more tickets by merging the second sub-sequence pattern with the first sub-sequence pattern based on the first score and the second score.

2. The method as claimed in claim 1 , wherein determining the POS weightage comprises:

assigning, by the issue identification system, a POS tag to each of one or more words in the one or more first sub-sequence patterns of the n-grams, to form a combination of the POS tags for each of the one or more first sub-sequence patterns; and

assigning, by the issue identification system, a predefined weightage for the combination of the POS tags for each of the one or more first sub-sequence patterns.

3. The method as claimed in claim 2 , wherein the predefined weightage is assigned based on a predefined priority associated with each of the one or more combinations of the POS tags.

4. The method as claimed in claim 1 , wherein the ticket data comprises ticket Identification (ID), description of one or more issues, the service category of the one or more tickets and other data related to the one or more tickets.

5. The method as claimed in claim 1 , wherein the one or more first sub-sequence patterns are generated by retaining order of words in the sequence pattern.

6. The method as claimed in claim 1 , wherein determining the POS weightage comprises:

assigning, by the issue identification system, a POS tag to each of one or more words of the one or more second sub-sequence patterns of the n-grams to form a combination of the POS tags for each of the one or more second sub-sequence patterns; and

assigning, by the issue identification system, a predefined weightage for the combination of the POS tags for each of the one or more second sub-sequence patterns.

7. The method as claimed in claim 6 , wherein the predefined weightage is assigned based on a predefined priority associated with each of the one or more combinations of the POS tags.

8. The method as claimed in claim 1 further comprises:

comparing, by the issue identification system, each of the one or more words in at least one of the first and the second sub-sequence patterns of the n-grams with one or more predefined domain keywords;

obtaining, by the issue identification system, at least one of the one or more first subsequence patterns and the one or more second sub-sequence patterns comprising the one or more predefined domain keywords as at least one of a representative first sub-sequence pattern and representative second sub-sequence pattern; and

identifying, by the issue identification system, automatically, one or more issues in each of the one or more tickets based on at least one of the representative first sub-sequence pattern and representative second sub-sequence pattern.

9. An issue identification system for automatically identifying one or more issues in one or more tickets of an organization, the issue identification system comprising:

a processor; and

a memory communicatively coupled to the processor, wherein the memory stores the processor-executable instructions, which, on execution, causes the processor to:

receive ticket data of one or more tickets related to a service category from one or more data sources;

generate one or more first sub-sequence patterns of n-grams for the one or more tickets from a sequence pattern retrieved from the ticket data;

determine frequency of occurrence of each of the one or more first sub-sequence patterns of the n-grams and a Part-of-Speech (POS) weightage of the one or more first sub-sequence patterns of the n-grams;

determine a first score for each of the one or more first sub-sequence patterns of the n-grams based on the frequency of occurrence and the POS weightage;

identify automatically, one or more issues in the one or more tickets based on the first sub-sequence pattern of the n-grams and the first score;

identify the sequence pattern corresponding to the one or more first sub-sequence patterns of the n-grams having the first score less than a predefined value for each of the one or more tickets;

generate one or more second sub-sequence patterns of the n-grams by removing one or more words in the sequence pattern in order of occurrence, wherein a distance value is associated with each of the one or more second sub-sequence patterns based on the one or more words removed in the sequence pattern;

determine a frequency of occurrence of each of the one or more second sub-sequence patterns of the n-grams and a POS weightage of the one or more second sub-sequence patterns of the n-grams;

determine a second score for each of the one or more second sub-sequence patterns of the n-grams based on the frequency of occurrence and the POS weightage of the one or more second sub-sequence patterns of the n-grams; and

update automatically, the one or more issues in the one or more tickets by merging the second sub-sequence pattern with the first sub-sequence pattern based on the first score and the second score.

10. The issue identification system as claimed in claim 9 , wherein the ticket data comprises ticket Identification (ID), description of one or more issues, the service category of the one or more tickets and other data related to the one or more tickets.

11. The issue identification system as claimed in claim 9 , wherein the processor is further configured to determine the POS weightage by:

assigning POS tags to each of one or more words of the one or more first sub-sequence patterns of the n-grams to form a combination of the POS tags for each of the one or more first sub-sequence patterns; and

assigning a predefined weightage for the combination of the POS tags for each of the one or more first sub-sequence patterns.

12. The issue identification system as claimed in claim 11 , wherein the processor assigns the predefined weightage based on a predefined priority associated with each of the one or more combinations of the POS tags.

13. The issue identification system as claimed in claim 9 , wherein the processor generates the one or more first sub-sequence patterns by retaining order of words in the sequence pattern.

14. The issue identification system as claimed in claim 9 , wherein the processor is further configured to determine the POS weightage by:

assigning POS tags to each of one or more words of the one or more second sub-sequence patterns of the n-grams to form a combination of the POS tags for each of the one or more second sub-sequence patterns; and

assigning a predefined weightage for the combination of the POS tags for each of the one or more second sub-sequence patterns.

15. The issue identification system as claimed in claim 14 , wherein the processor assigns the predefined weightage based on a predefined priority associated with each of the one or more combinations of the POS tags.

16. The issue identification system as claimed in claim 9 , wherein the processor is further configured to:

compare each of the one or more words in at least one of the first or the second sub-sequence patterns of the n-grams with one or more predefined domain keywords;

obtain at least one of the one or more first subsequence patterns and the one or more second sub-sequence patterns comprising the one or more predefined domain keywords as a pattern with the highest score based on the comparison; and

identify automatically, one or more issues in each of the one or more tickets based on the pattern with the highest score.

17. A non-transitory computer readable medium including instructions stored thereon that when processed by at least one processor causes an issue identification system to perform operations comprising:

receiving ticket data of one or more tickets related to a service category from one or more data sources;

generating one or more first sub-sequence patterns of n-grams for the one or more tickets from a sequence pattern retrieved from the ticket data;

determining a frequency of occurrence of each of the one or more first sub-sequence patterns of the n-grams and a Part-of-Speech (POS) weightage of the one or more first sub-sequence patterns of the n-grams;

determining a first score for each of the one or more first sub-sequence patterns of the n-grams based on the frequency of occurrence and the POS weightage;

identifying automatically, one or more issues in the one or more tickets based on the first sub-sequence pattern of the n-grams and the first score;

identifying the sequence pattern corresponding to the one or more first sub-sequence patterns of the n-grams having the first score less than a predefined value for each of the one or more tickets;

generating one or more second sub-sequence patterns of the n-grams by removing one or more words in the sequence pattern in order of occurrence, wherein a distance value is associated with each of the one or more second sub-sequence patterns based on the one or more words removed in the sequence pattern;

determining a frequency of occurrence of each of the one or more second sub-sequence patterns of the n-grams and a POS weightage of the one or more second sub-sequence patterns of the n-grams;

determining a second score for each of the one or more second sub-sequence patterns of the n-grams based on the frequency of occurrence and the POS weightage of the one or more second sub-sequence patterns of the n-grams; and

updating automatically, the one or more issues in the one or more tickets by merging the second sub-sequence pattern with the first sub-sequence pattern based on the first score and the second score.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 1, 2016
From: RAJARAM, VENKATAKRISHNAN; KONNAYAR, NARAYANAN RAMANI; CHAKRABORTY, RIA; SOUNDARARAJAN, MALATHI BELLAM
To: WIPRO LIMITED
Reel/Frame 038330/0334 →
Priority Claims (1)
IN 201641008636 · Mar 11, 2016 · national
Continuity (1)
Related Publication 20170262858A1 · Sep 14, 2017