IP Library Granted Patent US 11,798,549
Granted Patent B2
US 11,798,549 · App. 17/207,233 · Granted Oct 24, 2023

Generating action items during a conferencing session

Inventors: Jonathan Braganza (Ottawa, CA); Kevin Lee (Kanata, CA); Logendra Naidoo (Ottawa, CA)
Assignee: Mitel Networks Corporation
G10L15/22G10L15/10G10L15/1815H04L12/1831
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,798,549
App. No.
17/207,233
Granted
Oct 24, 2023
Kind
B2
Abstract

Embodiments include systems and methods for receiving an action item trigger by a user of a conferencing application; and in response to receiving the action item trigger, generating spoken words from audio data of a session of the conferencing application; normalizing the spoken words; generating higher-level representations of the normalized spoken words; determining semantic similarities of the higher-level representations of the normalized spoken words and higher level representations of normalized action words of an action word list; ranking options for top spoken words and action words based at least in part on the semantic similarities; identifying candidates for action words and/or phrases from the top spoken words and action words; and parsing the candidates to generate one or more action items.

Claims (45)

1. A method comprising:

receiving an action item trigger by a user of a conferencing application; and

in response to receiving the action item trigger:

generating spoken words from audio data of a session of the conferencing application for a previous predetermined number of sentences spoken before receiving the action item trigger;

normalizing the spoken words;

generating higher level representations of the normalized spoken words;

determining semantic similarities of the higher-level representations of the normalized spoken words and higher-level representations of normalized action words of an action word list;

ranking options for top spoken words and action words based at least in part on the semantic similarities;

identifying candidates for action words and/or phrases from the top spoken words and action words; and

parsing the candidates to generate one or more action items.

2. The method of claim 1 , comprising normalizing action words in the action word list and generating higher-level representations of the normalized action words.

3. The method of claim 1 , comprising parsing the candidates to generate one or more action items by:

dividing, by a trained neural network, candidate action phrases into words based at least in part on data from a dependency databank;

classifying the words into dependency classes according to a set of rules; and

generating one or more action items based at least in part on the dependency classes.

4. The method of claim 1 , wherein the action item trigger comprises an input selection in real-time by the user of a user interface of the conferencing application.

5. The method of claim 1 , wherein the one or more action items comprises a task assigned to be performed by a person based on analysis of the audio data of the session of the conferencing application.

6. The method of claim 5 , comprising sending an email to the person's email account notifying the person of the assigned task.

7. The method of claim 1 , comprising generating spoken words from audio data of the session of the conferencing application for a previous predetermined period of time before receiving the action item trigger instead of for the previous predetermined number of sentences spoken before receiving the action trigger.

8. The method of claim 1 , comprising excluding spoken words attributable to the user.

9. The method of claim 1 , wherein the action word list comprises a private lexicon.

10. At least one non-transitory machine-readable storage medium comprising instructions that, when executed, cause at least one processor to:

receive an action item trigger by a user of a conferencing application; and

in response to receiving the action item trigger:

generate spoken words from audio data of a session of the conferencing application for a previous predetermined number of sentences spoken before receiving the action item trigger;

normalize the spoken words;

generate higher-level representations of the normalized spoken words;

determine semantic similarities of the higher-level representations of the normalized spoken words and higher-level representations of normalized action words of an action word list;

rank options for top spoken words and action words based at least in part on the semantic similarities;

identify candidates for action words and/or phrases from the top spoken words and action words; and

parse the candidates to generate one or more action items.

11. The least one non-transitory machine-readable storage medium of claim 10 comprising instructions that, when executed, cause at least one processor to normalize action words in the action word list and generate higher level representations of the normalized action words.

12. The least one non-transitory machine-readable storage medium of claim 10 comprising instructions that, when executed, cause at least one processor to parse the candidates to generate one or more action items by:

dividing, by a trained neural network, candidate action phrases into words based at least in part on data from a dependency databank;

classifying the words into dependency classes according to a set of rules; and

generating one or more action items based at least in part on the dependency classes.

13. The least one non-transitory machine-readable storage medium of claim 10 , wherein the action item trigger comprises an input selection in real-time by the user of a user interface of the conferencing application.

14. The least one non-transitory machine-readable storage medium of claim 10 , wherein the one or more action items comprises a task assigned to be performed by a person based on analysis of the audio data of the session of the conferencing application.

15. An apparatus comprising:

a conferencing manager to receive audio data and an action item trigger; and

an action item generator to receive an action item trigger by a user of a conferencing application; and in response to receiving the action item trigger, to generate spoken words from the audio data of a session of the conferencing application for a previous predetermined number of sentences spoken before receiving the action item trigger; normalize the spoken words; generate higher level representations of the normalized spoken words; determine semantic similarities of the higher level representations of the normalized spoken words and higher level representations of normalized action words of an action word list; rank options for top spoken words and action words based at least in part on the semantic similarities; identify candidates for action words and/or phrases from the top spoken words and action words; and parse the candidates to generate one or more action items.

16. The apparatus of claim 15 wherein the action item generator is to normalize action words in the action word list and generate higher level representations of the normalized action words.

17. The apparatus of claim 15 wherein the action item generator to parse the candidates to generate one or more action items by dividing, by a trained neural network, candidate action phrases into words based at least in part on data from a dependency databank; classifying the words into dependency classes according to a set of rules; and generating one or more action items based at least in part on the dependency classes.

18. The apparatus of claim 15 , wherein the action item trigger comprises an input selection in real-time by the user of a user interface of the conferencing application.

19. The apparatus of claim 15 , wherein the one or more action items comprises a task assigned to be performed by a person based on analysis of the audio data of the session of the conferencing application.

Assignments (10)
SECURITY INTEREST Recorded Jun 30, 2025
From: MLN US HOLDCO LLC; MITEL (DELAWARE), INC.; MITEL NETWORKS CORPORATION; MITEL NETWORKS, INC.
To: U.S. PCI SERVICES, LLC
Reel/Frame 071758/0843 →
RELEASE OF SECURITY INTEREST Recorded Jun 24, 2025
From: WILMINGTON SAVINGS FUND SOCIETY, FSB
To: MITEL (DELAWARE), INC.; MITEL COMMUNICATIONS, INC.; MITEL NETWORKS, INC.; MITEL NETWORKS CORPORATION
Reel/Frame 071712/0821 →
RELEASE OF SECURITY INTEREST Recorded Jun 24, 2025
From: ACQUIOM AGENCY SERVICES LLC
To: MITEL (DELAWARE), INC.; MITEL NETWORKS, INC.; MITEL NETWORKS CORPORATION
Reel/Frame 071730/0632 →
SECURITY INTEREST Recorded Jun 20, 2025
From: MITEL (DELAWARE), INC.; MITEL NETWORKS CORPORATION; MITEL NETWORKS, INC.
To: ACQUIOM AGENCY SERVICES LLC
Reel/Frame 071676/0815 →
SECURITY INTEREST Recorded Mar 12, 2025
From: MITEL (DELAWARE), INC.; MITEL NETWORKS CORPORATION; MITEL NETWORKS, INC.
To: ACQUIOM AGENCY SERVICES LLC
Reel/Frame 070689/0857 →
NOTICE OF SUCCCESSION OF AGENCY - 2L Recorded Jan 14, 2025
From: UBS AG, STAMFORD BRANCH, AS LEGAL SUCCESSOR TO CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
To: WILMINGTON SAVINGS FUND SOCIETY, FSB
Reel/Frame 069896/0001 →
NOTICE OF SUCCCESSION OF AGENCY - 3L Recorded Jan 14, 2025
From: UBS AG, STAMFORD BRANCH, AS LEGAL SUCCESSOR TO CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
To: WILMINGTON SAVINGS FUND SOCIETY, FSB
Reel/Frame 070006/0268 →
NOTICE OF SUCCCESSION OF AGENCY - PL Recorded Jan 14, 2025
From: UBS AG, STAMFORD BRANCH, AS LEGAL SUCCESSOR TO CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
To: WILMINGTON SAVINGS FUND SOCIETY, FSB
Reel/Frame 069895/0755 →
SECURITY INTEREST Recorded Oct 31, 2022
From: MITEL NETWORKS CORPORATION
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
Reel/Frame 061824/0282 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 19, 2021
From: BRAGANZA, JONATHAN; LEE, KEVIN; NAIDOO, LOGENDRA
To: MITEL NETWORKS CORPORATION
Reel/Frame 055656/0057 →
Continuity (1)
Related Publication 20220301557A1 · Sep 22, 2022
Cited By (2)
US 12,598,270 US 12,641,193