Automated clinical documentation system and method
A method, computer program product, and computing system for synchronizing machine vision and audio is executed on a computing device and includes obtaining encounter information of a patient encounter, wherein the encounter information includes machine vision encounter information and audio encounter information. The machine vision encounter information and the audio encounter information are temporally-aligned to produce a temporarily-aligned encounter recording.
1. A computer-implemented method for synchronizing machine vision and audio, executed on a computing device, comprising:
obtaining encounter information of a patient encounter, wherein the encounter information includes machine vision encounter information and audio encounter information, wherein obtaining encounter information includes utilizing a virtual assistant to prompt an encounter participant to provide at least one of a pre-visit portion of the patient encounter and a post-visit portion of the patient encounter, wherein the virtual assistant verbally prompts the encounter participant to provide at least one of the pre-visit portion of the patient encounter and the post-visit portion of the patient encounter;
temporally aligning the machine vision encounter information and the audio encounter information to produce a temporally-aligned encounter recording;
receiving a user selection of at least a portion of the temporally-aligned encounter information to be rendered, wherein at least the portion of the temporally-aligned encounter recording selected by the user is the at least one of the pre-visit portion of the patient encounter obtained by the virtual assistant and the post-visit portion of the patient encounter obtained by the virtual assistant; and
rendering, visually, only the user selection of at least the portion of the temporally-aligned encounter recording selected by the user to include an indication of each of the portions of the audio encounter information during which the encounter participant of the patient encounter speaks in the pre-visit portion of the patient encounter when the pre-visit portion of the patient encounter is selected by the user, and rendering, visually, only the user selection of at least the portion of the temporally-aligned encounter recording selected by the user to include an indication of each of the portions of the audio encounter information during which the encounter participant of the patient encounter speaks in the post-visit portion of the patient encounter when the post-visit portion of the patient encounter is selected by the user.
2. The computer-implemented method of claim 1 wherein the machine vision encounter information is obtained via one or more machine vision systems.
3. The computer-implemented method of claim 2 wherein the one or more machine vision systems includes one or more of:
an RGB imaging system;
an infrared imaging system;
an ultraviolet imaging system;
a laser imaging system;
an X-ray imaging system;
a SONAR imaging system;
a RADAR imaging system; and
a thermal imaging system.
4. The computer-implemented method of claim 1 wherein the audio encounter information is obtained via one or more audio sensors.
5. The computer-implemented method of claim 1 wherein the indication of each of the portions of the audio encounter information during which the encounter participant of the patient encounter speaks is a visual indication of a display of each of the portions of the audio encounter information in at least the portion of the temporally-aligned encounter recording being rendered.
6. The computer-implemented method of claim 5 wherein the visual indication of each of the portions of the audio encounter information during which the encounter participant speaks is distinguished from the audio encounter information during which another encounter participant of the patient encounter speaks.
7. A computer program product residing on a non-transitory computer readable medium having a plurality of instructions stored thereon which, when executed by a processor, cause the processor to perform operations comprising:
obtaining encounter information of a patient encounter, wherein the encounter information includes machine vision encounter information and audio encounter information, wherein obtaining encounter information includes utilizing a virtual assistant to prompt an encounter participant to provide at least one of a pre-visit portion of the patient encounter and a post-visit portion of the patient encounter, wherein the virtual assistant verbally prompts the encounter participant to provide at least one of the pre-visit portion of the patient encounter and the post-visit portion of the patient encounter;
temporally aligning the machine vision encounter information and the audio encounter information to produce a temporally-aligned encounter recording;
receiving a user selection of at least a portion of the temporally-aligned encounter information to be rendered, wherein at least the portion of the temporally-aligned encounter recording selected by the user is the at least one of the pre-visit portion of the patient encounter obtained by the virtual assistant and the post-visit portion of the patient encounter obtained by the virtual assistant; and
rendering, visually, only the user selection of at least the portion of the temporally-aligned encounter recording selected by the user to include an indication of each of the portions of the audio encounter information during which the encounter participant of the patient encounter speaks in the pre-visit portion of the patient encounter when the pre-visit portion of the patient encounter is selected by the user, and rendering, visually, only the user selection of at least the portion of the temporally-aligned encounter recording selected by the user to include an indication of each of the portions of the audio encounter information during which the encounter participant of the patient encounter speaks in the post-visit portion of the patient encounter when the post-visit portion of the patient encounter is selected by the user.
8. The computer program product of claim 7 wherein the machine vision encounter information is obtained via one or more machine vision systems.
9. The computer program product of claim 7 wherein the one or more machine vision systems includes one or more of:
an RGB imaging system;
an infrared imaging system;
an ultraviolet imaging system;
a laser imaging system;
an X-ray imaging system;
a SONAR imaging system;
a RADAR imaging system; and
a thermal imaging system.
10. The computer program product of claim 7 wherein the audio encounter information is obtained via one or more audio sensors.
11. The computer program product of claim 7 wherein the indication of each of the portions of the audio encounter information during which the encounter participant of the patient encounter speaks is a visual indication of a display of each of the portions of the audio encounter information in at least the portion of the temporally-aligned encounter recording being rendered.
12. The computer program product of claim 11 wherein the visual indication of each of the portions of the audio encounter information during which the encounter participant speaks is distinguished from the audio encounter information during which another encounter participant of the patient encounter speaks.
13. A computing system including a processor and memory configured to perform operations comprising:
obtaining encounter information of a patient encounter, wherein the encounter information includes machine vision encounter information and audio encounter information, wherein obtaining encounter information includes utilizing a virtual assistant to prompt an encounter participant to provide at least one of a pre-visit portion of the patient encounter and a post-visit portion of the patient encounter, wherein the virtual assistant verbally prompts the encounter participant to provide at least one of the pre-visit portion of the patient encounter and the post-visit portion of the patient encounter;
temporally aligning the machine vision encounter information and the audio encounter information to produce a temporally-aligned encounter recording;
receiving a user selection of at least a portion of the temporally-aligned encounter information to be rendered, wherein at least the portion of the temporally-aligned encounter recording selected by the user is the at least one of the pre-visit portion of the patient encounter obtained by the virtual assistant and the post-visit portion of the patient encounter obtained by the virtual assistant; and
rendering, visually, only the user selection of at least the portion of the temporally-aligned encounter recording selected by the user to include an indication of each of the portions of the audio encounter information during which the encounter participant of the patient encounter speaks in the pre-visit portion of the patient encounter when the pre-visit portion of the patient encounter is selected by the user, and rendering, visually, only the user selection of at least the portion of the temporally-aligned encounter recording selected by the user to include an indication of each of the portions of the audio encounter information during which the encounter participant of the patient encounter speaks in the post-visit portion of the patient encounter when the post-visit portion of the patient encounter is selected by the user.
14. The computing system of claim 13 wherein the machine vision encounter information is obtained via one or more machine vision systems.
15. The computing system of claim 14 wherein the one or more machine vision systems includes one or more of:
an RGB imaging system;
an infrared imaging system;
an ultraviolet imaging system;
a laser imaging system;
an X-ray imaging system;
a SONAR imaging system;
a RADAR imaging system; and
a thermal imaging system.
16. The computing system of claim 13 wherein the audio encounter information is obtained via one or more audio sensors.
17. The computing system of claim 13 wherein the indication of each of the portions of the audio encounter information during which the encounter participant of the patient encounter speaks is a visual indication of a display of each of the portions of the audio encounter information in at least the portion of the temporally-aligned encounter recording being rendered.
18. The computing system of claim 17 wherein the visual indication of each of the portions of the audio encounter information during which the encounter participant speaks is distinguished from the audio encounter information during which another encounter participant of the patient encounter speaks.