IP Library Granted Patent US 12,683,014
Granted Patent B2
US 12,683,014 · App. 18/418,048 · Granted Jul 14, 2026

Machine-learning-oriented surgical video analysis system

Inventors: Jagadish Venkataraman (Menlo Park, CA); Pablo E. Garcia Kilroy (Menlo Park, CA)
Assignee: Auris Health, Inc.
G16H30/40G06N20/00G06V20/41G06V20/44G06V20/46G06V20/49G06V20/70G16H30/20G06V2201/03
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,683,014
App. No.
18/418,048
Filed
Jan 19, 2024
Granted
Jul 14, 2026
Kind
B2
Art Unit
3681
USPC
705/2
Abstract

Embodiments described herein provide various examples of a surgical video analysis system for segmenting surgical videos of a given surgical procedure into shorter video segments and labeling/tagging these video segments with multiple categories of machine learning descriptors. In one aspect, a process for processing surgical videos recorded during performed surgeries of a surgical procedure includes the steps of: receiving a diverse set of surgical videos associated with the surgical procedure; receiving a set of predefined phases for the surgical procedure and a set of machine learning descriptors identified for each predefined phase in the set of predefined phases; for each received surgical video, segmenting the surgical video into a set of video segments based on the set of predefined phases and for each segment of the surgical video of a given predefined phase, annotating the video segment with a corresponding set of machine learning descriptors for the given predefined phase.

Claims (47)

1 . A computer-implemented method, the method comprising:

receiving a surgical video of a surgical procedure performed by a surgeon as a plurality of video images;

identifying, in a plurality of subsets of video images of the plurality of video images, a surgical task, a surgical tool, and an anatomy;

for each subset of video images,

creating a metric that establishes an associative relationship among at least two of the surgical task, the surgical tool, or the anatomy within the subset of video images, and

annotating the subset of video images based on the metric, the surgical task, the surgical tool, or the anatomy;

computing an evaluation score for the surgeon according to an overall metric that comprises a union of all metrics from the plurality of subsets of video images; and

training a machine learning (ML) classifier using each annotated subset of video images to detect similar surgical tasks, surgical tools, anatomies, or metrics in other surgical videos.

2 . The computer-implemented method of claim 1 , wherein annotating comprises adding text or captions that describe the metric, the surgical task, the surgical tool, or the anatomy in the subset of video images based on 1) user input, 2) output of an annotation operation that is responsive to input based on the metric, the surgical task, the surgical tool, or the anatomy, or 3) a combination thereof.

3 . The computer-implemented method of claim 1 further comprising segmenting the surgical video into a plurality of phase segments corresponding to a set of predefined phases of the surgical procedure, wherein each phase segment comprises a different subset of video images of the plurality of video images and corresponds to a predefined phase of the set of predefined phases that is associated with the metric, the surgical task, the surgical tool, or the anatomy.

4 . The computer-implemented method of claim 1 further comprising producing a cleaned surgical video by identifying protected health information or non-intraoperative portions within one or more video images of the surgical video and removing the one or more video images from the surgical video.

5 . The computer-implemented method of claim 4 , wherein identifying comprises performing automatic object detection upon the cleaned surgical video.

6 . The computer-implemented method of claim 1 , wherein the ML classifier is a first ML classifier, wherein the method further comprises using the annotated subset of video images to train a second ML classifier to detect similar subsets of video images in other surgical videos.

7 . The computer-implemented method of claim 1 , wherein identifying comprises:

receiving a set of one or more ML descriptors; and

using the ML classifier to detect the surgical task, the surgical tool, or the anatomy within the subset of video images that match the set of one or more ML descriptors.

8 . A system comprising:

at least one processor; and

memory having stored instructions which when executed by the at least one processor causes the system to:

receive a surgical video of a surgical procedure performed by a surgeon as a plurality of video images;

identify, in a plurality of subsets of video images of the plurality of video images, a surgical task, a surgical tool, and an anatomy;

for each subset of video images,

creating a metric that establishes an associative relationship among at least two of the surgical task, the surgical tool, or the anatomy within the subset of video images, and

annotate the subset of video images based on the metric, the surgical task, the surgical tool, or the anatomy;

computing an evaluation score for the surgeon according to an overall metric that comprises a union of all metrics from the plurality of subsets of video images; and

train a machine learning (ML) classifier using each annotated subset of video images to detect similar surgical tasks, surgical tools, anatomies, or metrics in other surgical videos.

9 . The system of claim 8 , wherein the instructions to annotate comprises instructions to add text or captions that describe the metric, the surgical task, the surgical tool, or the anatomy in the subset of video images based on 1) user input, 2) output of an annotation operation that is responsive to input based on the metric, the surgical task, the surgical tool, or the anatomy, or 3) a combination thereof.

10 . The system of claim 8 , wherein the memory comprises further instructions to segment the surgical video into a plurality of phase segments corresponding to a set of predefined phases of the surgical procedure, wherein each phase segment is comprises a different subset of video images of the plurality of video images and corresponds to a predefined phase of the set of predefined phases that is associated with the metric, the surgical task, the surgical tool, or the anatomy.

11 . The system of claim 8 , wherein the memory has further instructions to produce a cleaned surgical video by identifying protected health information or non-intraoperative portions within one or more video images of the surgical video and removing the one or more video images from the surgical video.

12 . The system of claim 11 , wherein the instructions to identify comprises instructions to perform automatic object detection upon the cleaned surgical video.

13 . The system of claim 8 , wherein the ML classifier is a first ML classifier, wherein the memory comprises further instructions to use the annotated subset of video images to train a second ML classifier to detect similar subsets of video images in other surgical videos.

14 . The system of claim 8 , wherein instructions to identify comprises instructions to:

receive a set of one or more ML descriptors; and

use the ML classifier to detect the surgical task, the surgical tool, or the anatomy within the subset of video images that match the set of one or more ML descriptors.

15 . A non-transitory machine-readable medium comprising instructions which when executed by at least one processor of a system, causes the system to:

receive a surgical video of a surgical procedure performed by a surgeon as a plurality of video images;

identify, in a plurality of subsets of video images of the plurality of video images, a surgical task, a surgical tool, and an anatomy;

for each subset of video images,

creating a metric that establishes an associative relationship among at least two of the surgical task, the surgical tool, or the anatomy within the subset of video images, and

annotate the subset of video images based on the metric, the surgical task, the surgical tool, or the anatomy;

compute an evaluation score for the surgeon according to an overall metric that comprises a union of all metrics from the plurality of subsets of video images; and

train a machine learning (ML) classifier each annotated subset of video images to detect similar surgical tasks, surgical tools, anatomies, or metrics in other surgical videos.

16 . The non-transitory machine-readable medium of claim 15 , wherein the instructions to annotate comprises instructions to add text or captions that describe the metric, the surgical task, the surgical tool, or the anatomy in the subset of video images based on 1) user input, 2) output of an annotation operation that is responsive to input based on the metric, the surgical task, the surgical tool, or the anatomy, or 3) a combination thereof.

17 . The non-transitory machine-readable medium of claim 15 comprises further instructions to segment the surgical video into a plurality of phase segments corresponding to a set of predefined phases of the surgical procedure, wherein each phase segment comprises a different subset of video images of the plurality of video images and corresponds to a predefined phase of the set of predefined phases that is associated with the metric, the surgical task, the surgical tool, or the anatomy.

18 . The non-transitory machine-readable medium of claim 15 comprises further instructions to produce a cleaned surgical video by identifying protected health information or non-intraoperative portions within one or more video images of the surgical video and removing the one or more video images from the surgical video.

19 . The non-transitory machine-readable medium of claim 18 , wherein the instructions to identify comprises instructions to perform automatic object detection upon the cleaned surgical video.

20 . The non-transitory machine-readable medium of claim 15 , wherein the ML classifier is a first ML classifier, wherein the non-transitory machine-readable medium comprises further instructions to use the annotated subset of video images to train a second ML classifier to detect similar subsets of video images in other surgical videos.

Assignments (1)
MERGER Recorded Jan 27, 2026
From: VERB SURGICAL INC.
To: AURIS HEALTH, INC.
Reel/Frame 073601/0736 →
Continuity (3)
Continuation 17530232 · Nov 18, 2021
Continuation 15987782 · May 23, 2018
Related Publication 20240242818A1 · Jul 18, 2024
References Cited (83)
US 5740801A · Branson · 1998 [cited by examiner]
US 6920347B2 · Simon · 2005 [cited by examiner]
US 7068842B2 · Liang · 2006 [cited by examiner]
US 7317955B2 · McGreevy · 2008 [cited by examiner]
US 7379790B2 · Toth et al. · 2008 [cited by applicant]
US 7840042B2 · Kriveshko · 2010 [cited by examiner]
US 7853305B2 · Simon et al. · 2010 [cited by applicant]
US 8086008B2 · Coste-Maniere · 2011 [cited by examiner]
US 8108072B2 · Zhao · 2012 [cited by examiner]
US 8131031B2 · Lloyd · 2012 [cited by examiner]
US 8147503B2 · Zhao · 2012 [cited by examiner]
US 8443279B1 · Hameed · 2013 [cited by examiner]
US 8504136B1 · Sun · 2013 [cited by examiner]
US 8527094B2 · Kumar · 2013 [cited by examiner]
US 8600551B2 · Itkowitz · 2013 [cited by examiner]
US 8706184B2 · Mohr · 2014 [cited by examiner]
US 9025247B1 · Mossberg et al. · 2015 [cited by applicant]
US 9215293B2 · Miller · 2015 [cited by examiner]
US 9413976B2 · DiCarlo · 2016 [cited by examiner]
US 10169535B2 · Mentis · 2019 [cited by examiner]
US 10433914B2 · Wollowick · 2019 [cited by examiner]
US 10499996B2 · de Almeida Barreto · 2019 [cited by examiner]
US 10588699B2 · Richmond · 2020 [cited by examiner]
US 10679743B2 · Venkataraman · 2020 [cited by examiner]
US 10740552B2 · Hanning · 2020 [cited by examiner]
US 10803320B2 · Calmus · 2020 [cited by examiner]
US 11081229B2 · Alvi · 2021 [cited by examiner]
US 11176945B2 · Paul · 2021 [cited by examiner]
US 11189379B2 · Giataganas · 2021 [cited by examiner]
US 11202676B2 · Lightcap · 2021 [cited by examiner]
US 11205508B2 · Venkataraman · 2021 [cited by examiner]
US 20030208196A1 · Stone · 2003 [cited by examiner]
US 20050251156A1 · Toth · 2005 [cited by examiner]
US 20080003555A1 · Ekvall et al. · 2008 [cited by applicant]
US 20100285438A1 · Kesavadas · 2010 [cited by examiner]
US 20110301447A1 · Park · 2011 [cited by examiner]
US 20120046659A1 · Mueller · 2012 [cited by applicant]
US 20120253360A1 · White · 2012 [cited by examiner]
US 20130211588A1 · Diolaiti · 2013 [cited by applicant]
US 20140286533A1 · Luo · 2014 [cited by examiner]
US 20150005622A1 · Zhao · 2015 [cited by examiner]
US 20150230875A1 · Shademan · 2015 [cited by examiner]
US 20160100909A1 · Wollowick · 2016 [cited by examiner]
US 20160103810A1 · Hanning · 2016 [cited by examiner]
US 20160140875A1 · Kumar et al. · 2016 [cited by applicant]
US 20160166345A1 · Kumar et al. · 2016 [cited by applicant]
US 20160210411A1 · Mentis · 2016 [cited by examiner]
US 20170035517A1 · Geri et al. · 2017 [cited by applicant]
US 20170132785A1 · Wshah · 2017 [cited by examiner]
US 20180174311A1 · Kluckner et al. · 2018 [cited by applicant]
US 20180357514A1 · Zisimopoulos · 2018 [cited by examiner]
US 20190069957A1 · Barral · 2019 [cited by examiner]
US 20190362834A1 · Venkataraman · 2019 [cited by examiner]
US 20210000461A1 · Charles · 2021 [cited by examiner]
US 20210290317A1 · Sen · 2021 [cited by examiner]
US 20240242818A1 · Venkataraman · 2024 [cited by examiner]
US 20250062020A1 · Gordon · 2025 [cited by examiner]
CN 104000655A · 2014 [cited by applicant]
CN 105992996A · 2016 [cited by applicant]
CN 107667380A · 2018 [cited by applicant]
EP 2420197A2 · 2012 [cited by applicant]
KR 1020080001622A · 2008 [cited by applicant]
KR 1020140126322A · 2014 [cited by applicant]
WO 2016200887A1 · 2016 [cited by applicant]
WO 2017075541A1 · 2017 [cited by applicant]
Advisory Action received for U.S. Appl. No. 15/987,782, mailed on Nov. 10, 2020, 2 pages. [cited by applicant]
Extended European Search Report for European Application No. 18920011.6 mailed Feb. 3, 2022, 10 pages. [cited by applicant]
Extended European Search Report for European Application No. 18933279.4 mailed May 17, 2022, 9 pages. [cited by applicant]
Final Office Action received for U.S. Appl. No. 15/987,782, mailed on Aug. 17, 2020, 24 pages. [cited by applicant]
International Preliminary Report on Patentability for International Application No. PCT/US2018/036452 mailed Dec. 3, 2020, 7 pages. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/US2018/036452 mailed Aug. 31, 2018, 8 pages. [cited by applicant]
Lin, Henry C., et al., “Towards automatic skill evaluation: Detection and segmentation of robot-assisted surgical motions,” Computer Aided Surgery, vol. 11, No. 5, Dec. 31, 2006, pp. 220-230. [cited by applicant]
Loukas, C. Video content analysis of surgical procedures. Surg Endosc 32, 553-568 (2018). https://doi.org/10.1007/s00464-017-5878-1. Received: Feb. 14, 2017 /Accepted: Sep. 7, 2017 / Published online: Oct. 26, 2017 © Sp… [cited by applicant]
Non-Final Office Action received for U.S. Appl. No. 15/987,782, mailed on Feb. 5, 2021, 30 pages. [cited by applicant]
Non-Final Office Action received for U.S. Appl. No. 15/987,782, mailed on Jan. 24, 2020, 20 pages. [cited by applicant]
Non-Final Office Action received for U.S. Appl. No. 17/530,232, mailed on Feb. 17, 2023, 20 pages. [cited by applicant]
Notice of Allowance for U.S. Appl. No. 16/894,018 mailed Sep. 29, 2022, 10 pages. [cited by applicant]
Notice of Allowance received for U.S. Appl. No. 15/987,782, mailed on Aug. 18, 2021, 11 pages. [cited by applicant]
Notification of Reasons for Refusal for Japanese Application No. 2020-562174 mailed Jan. 18, 2022, 6 pages. [cited by applicant]
Office Action received for Chinese Patent Application No. 201880001594.4, mailed on Feb. 1, 2024, 24 pages (14 pages of English Translation and 10 pages of Original Document). [cited by applicant]
Office Action received for European Application No. 18920011.6, mailed on Jan. 24, 2024, 5 pages. [cited by applicant]
Padoy, Nicolas, “Workflow and Activity Modeling for Monitoring Surgical Procedures,” HAL archvies-ouvertes.fr, retrieved from the Internet <http://www.theses.fr/2010NAN10025/document, Apr. 14, 2010, 166 pages. [cited by applicant]
European Search Report received for European Patent Application No. 25205240.2, mailed Dec. 10, 2025, 8 pages. [cited by applicant]