IP Library Granted Patent US 11,605,448
Granted Patent B2
US 11,605,448 · App. 16/058,914 · Granted Mar 14, 2023

Automated clinical documentation system and method

Inventors: Donald E. Owen (Orlando, FL); Garret N. Erskine (Torrance, CA); Mehmet Mert Öz (Baden, AT); Daniel Paulino Almendro Barreda (London, GB)
Assignee: Nuance Communications, Inc.
G16H10/60A61B5/7405G06F3/16G06F16/637G06F16/685G06F16/904G06F21/6245G06F40/174G06F40/30G06F40/40G06K9/6288G06K19/07762G06N3/006G06T7/00G06V20/10G06V20/52G06V40/103G06V40/16G06V40/172G06V40/23G10L15/08G10L17/00G10L21/0232G11B27/10G16B50/00G16H10/20G16H15/00G16H30/00G16H30/20G16H30/40G16H40/20G16H40/60G16H40/63G16H50/20G16H50/30G16H80/00G16Y20/00H04L51/02H04L51/222H04N7/183H04R1/326H04R3/005H04R3/12G06T2207/10024G06T2207/10044G06T2207/10048G06T2207/10116G06T2207/10132G10L15/1815G10L15/22G10L15/26G10L2021/02082H04N7/181H04R1/406H04R3/02H04R2420/07H04S2400/15
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,605,448
App. No.
16/058,914
Granted
Mar 14, 2023
Kind
B2
Abstract

A method, computer program product, and computing system for visual diarization of an encounter is executed on a computing device and includes obtaining encounter information of a patient encounter. The encounter information is processed to: associate a first portion of the encounter information with a first encounter participant, and associate at least a second portion of the encounter information with at least a second encounter participant. A visual representation of the encounter information is rendered. A first visual representation of the first portion of the encounter information is rendered that is temporally-aligned with the visual representation of the encounter information. At least a second visual representation of the at least a second portion of the encounter information is rendered that is temporally-aligned with the visual representation of the encounter information.

Claims (57)

1. A computer-implemented method for visual diarization of an encounter, executed on a computing device, comprising:

obtaining encounter information of a patient encounter, wherein obtaining the encounter information includes one of utilizing a virtual assistant to audibly prompt a patient to provide at least a portion of the encounter information during a pre-visit portion of the patient encounter and utilizing the virtual assistant to audibly prompt the patient to provide at least a portion of the encounter information during a post-visit portion of the patient encounter, wherein the encounter information includes at least one of audio encounter information obtained from the virtual assistant via one or more audio sensors and machine vision encounter information obtained from the virtual assistant via one or more machine vision systems;

processing the encounter information to:

associate a first portion of the encounter information with a first encounter participant, wherein the first portion of the encounter information includes a first portion of the at least one of audio encounter information obtained via the one or more audio sensors and the machine vision encounter information obtained via the one or more machine vision systems, and wherein the first encounter participant is a medical professional, and

associate at least a second portion of the encounter information with at least a second encounter participant, wherein the second portion of the encounter information includes a second portion of the at least one of audio encounter information obtained from the virtual assistant via the one or more audio sensors and the machine vision encounter information obtained from the virtual assistant via the one or more machine vision systems, wherein at least the second encounter participant is the patient;

rendering a visual representation of the encounter information obtained via the one or more audio sensors and the machine vision encounter information obtained via the one or more machine vision systems, wherein the visual representation of the encounter information includes the machine vision encounter information;

rendering a first visual representation of the first portion of the encounter information that is temporally-aligned with the visual representation of the encounter information, wherein rendering of the first visual representation is filtered based upon, at least in part, a selection on a user interface with playback controls displaying the encounter information by a user of the first portion of the at least one of audio encounter information obtained via the one or more audio sensors and the machine vision encounter information obtained via the one or more machine vision systems; and

rendering at least a second visual representation of the at least the second portion of the encounter information that is temporally-aligned with the visual representation of the encounter information, wherein rendering of the at least the second visual representation is filtered based upon, at least in part, a selection on the user interface with the playback controls displaying the encounter information by the user, wherein the selection by the user on the user interface with playback controls is a selection of the second visual representation of the second portion of the at least one of audio encounter information obtained from the virtual assistant via the one or more audio sensors and the machine vision encounter information obtained from the virtual assistant via the one or more machine vision systems, wherein filtering includes visually annotating the visual representation of one of the first and second visual representations indicating whether the encounter information in the encounter transcript is one of the pre-visit portion of the patient encounter and the post-visit portion of the patient encounter.

2. The computer-implemented method of claim 1 further comprising:

generating an encounter transcript based, at least in part, upon the first portion of the encounter information and the at least the second portion of the encounter information.

3. The computer-implemented method of claim 1 wherein the one or more machine vision systems includes one or more of:

an RGB imaging system;

an infrared imaging system;

an ultraviolet imaging system;

a laser imaging system;

an X-ray imaging system;

a SONAR imaging system;

a RADAR imaging system; and

a thermal imaging system.

4. A computer program product residing on a non-transitory computer readable medium having a plurality of instructions stored thereon which, when executed by a processor, cause the processor to perform operations comprising:

obtaining encounter information of a patient encounter, wherein obtaining the encounter information includes one of utilizing a virtual assistant to audibly prompt a patient to provide at least a portion of the encounter information during a pre-visit portion of the patient encounter and utilizing the virtual assistant to audibly prompt the patient to provide at least a portion of the encounter information during a post-visit portion of the patient encounter, wherein the encounter information includes at least one of audio encounter information obtained from the virtual assistant via one or more audio sensors and machine vision encounter information obtained from the virtual assistant via one or more machine vision systems;

processing the encounter information to:

associate a first portion of the encounter information with a first encounter participant, wherein the first portion of the encounter information includes a first portion of the at least one of audio encounter information obtained via the one or more audio sensors and the machine vision encounter information obtained via the one or more machine vision systems, and wherein the first encounter participant is a medical professional, and

associate at least a second portion of the encounter information with at least a second encounter participant, wherein the second portion of the encounter information includes a second portion of the at least one of audio encounter information obtained from the virtual assistant via the one or more audio sensors and the machine vision encounter information obtained from the virtual assistant via the one or more machine vision systems, wherein at least the second encounter participant is the patient;

rendering a visual representation of the encounter information obtained via the one or more audio sensors and the machine vision encounter information obtained via the one or more machine vision systems, wherein the visual representation of the encounter information includes the machine vision encounter information;

rendering a first visual representation of the first portion of the encounter information that is temporally-aligned with the visual representation of the encounter information, wherein rendering of the first visual representation is filtered based upon, at least in part, a selection on a user interface with playback controls displaying the encounter information by a user of the first portion of the at least one of audio encounter information obtained via the one or more audio sensors and the machine vision encounter information obtained via the one or more machine vision systems; and

rendering at least a second visual representation of the at least the second portion of the encounter information that is temporally-aligned with the visual representation of the encounter information, wherein rendering of the at least the second visual representation is filtered based upon, at least in part, a selection on the user interface with the playback controls displaying the encounter information by the user, wherein the selection by the user on the user interface with playback controls is a selection of the second visual representation of the second portion of the at least one of audio encounter information obtained from the virtual assistant via the one or more audio sensors and the machine vision encounter information obtained from the virtual assistant via the one or more machine vision systems, wherein filtering includes visually annotating the visual representation of one of the first and second visual representations indicating whether the encounter information in the encounter transcript is one of the pre-visit portion of the patient encounter and the post-visit portion of the patient encounter.

5. The computer program product of claim 4 further comprising:

generating an encounter transcript based, at least in part, upon the first portion of the encounter information and the at least the second portion of the encounter information.

6. The computer program product of claim 4 wherein the one or more machine vision systems includes one or more of:

an RGB imaging system;

an infrared imaging system;

an ultraviolet imaging system;

a laser imaging system;

an X-ray imaging system;

a SONAR imaging system;

a RADAR imaging system; and

a thermal imaging system.

7. A computing system including a processor and memory configured to perform operations comprising:

obtaining encounter information of a patient encounter, wherein obtaining the encounter information includes one of utilizing a virtual assistant to audibly prompt a patient to provide at least a portion of the encounter information during a pre-visit portion of the patient encounter and utilizing the virtual assistant to audibly prompt the patient to provide at least a portion of the encounter information during a post-visit portion of the patient encounter, wherein the encounter information includes at least one of audio encounter information obtained from the virtual assistant via one or more audio sensors and machine vision encounter information obtained from the virtual assistant via one or more machine vision systems;

processing the encounter information to:

associate a first portion of the encounter information with a first encounter participant, wherein the first portion of the encounter information includes a first portion of the at least one of audio encounter information obtained via the one or more audio sensors and the machine vision encounter information obtained via the one or more machine vision systems, and wherein the first encounter participant is a medical professional, and

associate at least a second portion of the encounter information with at least a second encounter participant, wherein the second portion of the encounter information includes a second portion of the at least one of audio encounter information obtained from the virtual assistant via the one or more audio sensors and the machine vision encounter information obtained from the virtual assistant via the one or more machine vision systems, wherein at least the second encounter participant is the patient;

rendering a visual representation of the encounter information obtained via the one or more audio sensors and the machine vision encounter information obtained via the one or more machine vision systems, wherein the visual representation of the encounter information includes the machine vision encounter information;

rendering a first visual representation of the first portion of the encounter information that is temporally-aligned with the visual representation of the encounter information, wherein rendering of the first visual representation is filtered based upon, at least in part, a selection on a user interface with playback controls displaying the encounter information by a user of the first portion of the at least one of audio encounter information obtained via the one or more audio sensors and the machine vision encounter information obtained via the one or more machine vision systems; and

rendering at least a second visual representation of the at least the second portion of the encounter information that is temporally-aligned with the visual representation of the encounter information, wherein rendering of the at least the second visual representation is filtered based upon, at least in part, a selection on the user interface with the playback controls displaying the encounter information by the user, wherein the selection by the user on the user interface with playback controls is a selection of the second visual representation of the second portion of the at least one of audio encounter information obtained from the virtual assistant via the one or more audio sensors and the machine vision encounter information obtained from the virtual assistant via the one or more machine vision systems, wherein filtering includes visually annotating the visual representation of one of the first and second visual representations indicating whether the encounter information in the encounter transcript is one of the pre-visit portion of the patient encounter and the post-visit portion of the patient encounter.

8. The computing system of claim 7 further comprising:

generating an encounter transcript based, at least in part, upon the first portion of the encounter information and the at least the second portion of the encounter information.

9. The computing system of claim 7 wherein the one or more machine vision systems includes one or more of:

an RGB imaging system;

an infrared imaging system;

an ultraviolet imaging system;

a laser imaging system;

an X-ray imaging system;

a SONAR imaging system;

a RADAR imaging system; and

a thermal imaging system.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 9, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065532/0152 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 13, 2018
From: OWEN, DONALD E; ERSKINE, GARRET N; ÖZ, MEHMET MERT; BARREDA, DANIEL PAULINO ALMENDRO
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 047068/0950 →
Cited By (1)
US 12,236,323