IP Library › Granted Patent US 12,089,820
Granted Patent B2
US 12,089,820 · App. 17/251,768 · Granted Sep 17, 2024

Systems and methods for processing real-time video from a medical image device and detecting objects in the video

Inventors: Nhan Ngo Dinh (Rome, IT); Giulio Evangelisti (Rome, IT); Flavio Navari (Rome, IT)
Assignee: COSMO ARTIFICIAL INTELLIGENCE—AI LIMITED
A61B1/31A61B1/000094A61B1/000095A61B1/000096A61B1/00055A61B5/7264A61B5/7267G06F18/214G06N3/045G06N3/08G06T7/0012G06T7/70G06T11/001G06T11/203G06T11/60G06V10/25G06V10/255G06V10/82G06V20/40G06V20/49G16H30/20G06T2207/10016G06T2207/10068G06T2207/20084G06T2207/30004G06T2207/30032G06T2207/30064G06T2207/30096G06V2201/03G06V2201/032
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,089,820
App. No.
17/251,768
Granted
Sep 17, 2024
Kind
B2
Abstract

The present disclosure relates to systems and methods for processing real-time video and detecting objects in the video. In one implementation, a system is provided that includes an input port for receiving real-time video obtained from a medical image device, a first bus for transferring the received real-time video, and at least one processor configured to receive the real-time video from the first bus, perform object detection by applying a trained neural network on frames of the received real-time video, and overlay a border indicating a location of at least one detected object in the frames. The system also includes a second bus for receiving the video with the overlaid border, an output port for outputting the video with the overlaid border from the second bus to an external display, and a third bus for directly transmitting the received real-time video to the output port.

Claims (38)

1. A computer-implemented system for real-time video processing, comprising:

a medical imaging device configured to capture real-time video during a medical procedure, the real-time video comprising a plurality of frames;

one or more neural networks that implement:

a trained object detector configured to directly receive the plurality of frames captured by the medical imaging device as input and to generate at least one detection of a feature-of-interest in the plurality of frames as output; and

a trained classifier configured to receive the at least one detection of the feature-of-interest from the trained object detector as input and to generate, based on at least one category, a classification of the feature-of-interest as output;

a display device configured to display to an operator of the medical imaging device, in real-time, the plurality of frames together with an overlay based on the at least one detection of the feature-of-interest and the classification of the feature-of-interest;

an input device configured to receive from the operator a command that adjusts a sensitivity setting associated with at least one of the detection and the classification; and

at least one processor configured to adjust, during the medical procedure in response to the command, one or more weights of one or more nodes of at least one of the trained object detector and the trained classifier.

2. The system of claim 1 , wherein the medical procedure is at least one of an endoscopy, a gastroscopy, a colonoscopy, and an enteroscopy.

3. The system of claim 1 , wherein the feature-of-interest includes an abnormality of human tissue comprising at least one of a formation on or of the human tissue, a change in the human tissue from one type of cell to another type of cell, an absence of the human tissue from a location where the human tissue is expected, and a lesion.

4. The system of claim 1 , wherein the at least one processor is further configured to adjust, in response to the command, one or more thresholds associated with an output layer of at least one of the trained object detector and the trained classifier.

5. The system of claim 1 , wherein the overlay indicates a location of the at least one detection of the feature-of-interest in the plurality of frames.

6. The system of claim 5 , wherein the overlay comprises a bounding box.

7. The system of claim 5 , wherein the display of the overlay is modified based on the classification generated by the trained classifier.

8. The system of claim 7 , wherein the display of the overlay is modified to a first color if the trained classifier classifies the feature-of-interest in a first category, and is modified to a second color if the trained classifier classifies the feature-of-interest in a second category.

9. The system of claim 1 , wherein the classification is based on at least one of a histological classification, morphological classification, and a structural classification.

10. The system of claim 1 , wherein one or more weights of one or more nodes of the trained object detector is adjusted based on the sensitivity setting; and wherein a number of detections that are produced by the trained object detector is increased or decreased in accordance with the sensitivity setting.

11. A method for real-time video processing, the method comprising:

capturing, via a medical imaging device, real-time video during a medical procedure, the real-time video comprising a plurality of frames;

directly providing the real-time video comprising the plurality of frames to a trained object detector, the trained object detector comprising one or more neural networks;

generating at least one detection of a feature-of-interest by applying the trained object detector to the plurality of frames;

generating, based on at least one category, a classification of the feature-of-interest by applying a trained classifier to the at least one detection, the trained classifier comprising one or more neural networks;

displaying to an operator of the medical imaging device, via a display device, the plurality of frames together with an overlay based on the at least one detection of the feature-of-interest and the classification of the feature-of-interest;

receiving from the operator, via an input device, a command that adjusts a sensitivity setting associated with at least one of the detection and the classification; and

adjusting, during the medical procedure in response to the command, one or more weights of one or more nodes of at least one of the trained object detector and the trained classifier.

12. The method of claim 11 , wherein the medical procedure is at least one of an endoscopy, a gastroscopy, a colonoscopy, and an enteroscopy.

13. The method of claim 11 , wherein the feature-of-interest includes an abnormality of human tissue comprising at least one of a formation on or of the human tissue, a change in the human tissue from one type of cell to another type of cell, an absence of the human tissue from a location where the human tissue is expected, and a lesion.

14. The method of claim 11 , further comprising:

adjusting, in response to the command, one or more thresholds associated with an output layer of at least one of the trained object detector and the trained classifier.

15. The method of claim 11 , wherein the overlay indicates a location of the at least one detection of the feature-of-interest in the plurality of frames.

16. The method of claim 15 , wherein the overlay comprises a bounding box.

17. The method of claim 15 , further comprising:

modifying the display of the overlay based on the classification generated by the trained classifier.

18. The method of claim 17 , further comprising:

modifying the display of the overlay to a first color if the trained classifier classifies the feature-of-interest in a first category; and

modifying the overlay to a second color if the trained classifier classifies the feature-of-interest in a second category.

19. The method of claim 11 , wherein the classification is based on at least one of a histological classification, morphological classification, and a structural classification.

20. The method of claim 11 , wherein one or more weights of one or more nodes of the trained object detector is adjusted based on the sensitivity setting; and wherein a number of detections that are produced by the trained object detector is increased or decreased in accordance with the sensitivity setting.

Assignments (5)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 12, 2024
From: NGO DINH, NHAN; EVANGELISTI, GIULIO; NAVARI, FLAVIO
To: COSMO ARTIFICIAL INTELLIGENCE – AI LIMITED
Reel/Frame 067972/0944 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 30, 2024
From: LINKVERSE S.R.L.
To: COSMO TECHNOLOGIES LIMITED
Reel/Frame 067276/0180 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 30, 2024
From: GRANELL STRATEGIC INVESTMENT FUND LIMITED
To: COSMO ARTIFICIAL INTELLIGENCE - AI LIMITED
Reel/Frame 067276/0189 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 30, 2024
From: COSMO TECHNOLOGIES LIMITED
To: COSMO ARTIFICIAL INTELLIGENCE - AI LIMITED
Reel/Frame 067276/0191 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 30, 2024
From: NAVARI, FLAVIO; NGO DINH, NHAN; EVANGELISTI, GIULIO
To: LINKVERSE S.R.L.
Reel/Frame 067276/0207 →
Priority Claims (1)
EP 18180572 · Jun 28, 2018 · regional
Continuity (1)
Related Publication 20210133972A1 · May 6, 2021