IP Library Granted Patent US 11,705,232
Granted Patent B2
US 11,705,232 · App. 17/669,099 · Granted Jul 18, 2023

Communication system and method

Inventor: Joel Praveen Pinto (Aachen, DE)
Assignee: Nuance Communications, Inc.
G16H15/00G06F3/013G06F40/186G06F40/40G10L15/26
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,705,232
App. No.
17/669,099
Granted
Jul 18, 2023
Kind
B2
Abstract

A method, computer program product, and computing system for receiving audio-based content from a user who is reviewing an image on a display screen; receiving gaze information that defines a gaze location of the user; and temporally aligning the audio-based content and the gaze information to form location-based content.

Claims (42)

1. A computer-implemented method, executed on a computing device, comprising:

receiving audio-based content from a user who is reviewing an image on a display screen;

receiving gaze information that defines a first gaze location and a second gaze location of the user, wherein the gaze information defines a first gaze location of the user with respect to a first portion of the image on the display screen and a second gaze location of the user with respect to a second portion of the image on the display screen;

temporally aligning the audio-based content and the gaze information to form location-based content associated with each of the first portion of the image on the display screen and the second portion of the image on the display screen; and

associating the audio-based content from the user with each respective first and second portion of the image on the display screen defined by each respective first and second gaze location of the user.

2. The computer-implemented method of claim 1 further comprising:

rendering the image on the display screen.

3. The computer-implemented method of claim 1 wherein the audio-based content and the gaze information are generated during a telehealth session.

4. The computer-implemented method of claim 1 wherein the audio-based content and the gaze information are generated during a diagnostic session.

5. The computer-implemented method of claim 1 wherein the audio-based content and the gaze information are generated during an AI model training session.

6. The computer-implemented method of claim 1 wherein the location-based content is utilized to populate a medical report.

7. The computer-implemented method of claim 1 wherein the location-based content is utilized to train an AI model.

8. The computer-implemented method of claim 1 wherein the location-based content is utilized to enrich a medical image.

9. The computer-implemented method of claim 8 wherein the audio-based content is descriptive content that describes a portion of the medical image.

10. A computer program product residing on a computer readable medium having a plurality of instructions stored thereon which, when executed by a processor, cause the processor to perform operations comprising:

receiving audio-based content from a user who is reviewing an image on a display screen;

receiving gaze information that defines a first gaze location and a second gaze location of the user, wherein the gaze information defines a first gaze location of the user with respect to a first portion of the image on the display screen and a second gaze location of the user with respect to a second portion of the image on the display screen;

temporally aligning the audio-based content and the gaze information to form location-based content associated with each of the first portion of the image on the display screen and the second portion of the image on the display screen; and

associating the audio-based content from the user with each respective first and second portion of the image on the display screen defined by each respective first and second gaze location of the user.

11. The computer-implemented method of claim 10 further comprising:

rendering the image on the display screen.

12. The computer-implemented method of claim 10 wherein the audio-based content and the gaze information are generated during a telehealth session.

13. The computer-implemented method of claim 10 wherein the audio-based content and the gaze information are generated during a diagnostic session.

14. The computer-implemented method of claim 10 wherein the audio-based content and the gaze information are generated during an AI model training session.

15. The computer-implemented method of claim 10 wherein the location-based content is utilized to populate a medical report.

16. The computer-implemented method of claim 10 wherein the location-based content is utilized to train an AI model.

17. The computer-implemented method of claim 10 wherein the location-based content is utilized to enrich a medical image.

18. The computer-implemented method of claim 17 wherein the audio-based content is descriptive content that describes a portion of the medical image.

19. A computing system including a processor and memory configured to perform operations comprising:

receiving audio-based content from a user who is reviewing an image on a display screen;

receiving gaze information that defines a first gaze location and a second gaze location of the user, wherein the gaze information defines a first gaze location of the user with respect to a first portion of the image on the display screen and a second gaze location of the user with respect to a second portion of the image on the display screen;

temporally aligning the audio-based content and the gaze information to form location-based content associated with each of the first portion of the image on the display screen and the second portion of the image on the display screen; and

associating the audio-based content from the user with each respective first and second portion of the image on the display screen defined by each respective first and second gaze location of the user.

20. The computer-implemented method of claim 19 further comprising:

rendering the image on the display screen.

21. The computer-implemented method of claim 19 wherein the audio-based content and the gaze information are generated during a telehealth session.

22. The computer-implemented method of claim 19 wherein the audio-based content and the gaze information are generated during a diagnostic session.

23. The computer-implemented method of claim 19 wherein the audio-based content and the gaze information are generated during an AI model training session.

24. The computer-implemented method of claim 19 wherein the location-based content is utilized to populate a medical report.

25. The computer-implemented method of claim 19 wherein the location-based content is utilized to train an AI model.

26. The computer-implemented method of claim 19 wherein the location-based content is utilized to enrich a medical image.

27. The computer-implemented method of claim 26 wherein the audio-based content is descriptive content that describes a portion of the medical image.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 14, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065578/0676 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 10, 2022
From: PINTO, JOEL PRAVEEN
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 058976/0913 →