IP Library › Granted Patent US 12,266,098
Granted Patent B2
US 12,266,098 · App. 17/523,937 · Granted Apr 1, 2025

Improving model performance by artificial blending of healthy tissue

Inventors: Yoel Shoshan (Haifa, IL); Vadim Ratner (Haifa, IL)
Assignee: International Business Machines Corporation
G06T7/0012G06F18/2431G06N3/08G06T2207/20081G06T2207/20084G06T2207/30024
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,266,098
App. No.
17/523,937
Granted
Apr 1, 2025
Kind
B2
Abstract

A method for improving machine learning algorithm performance is described. The method may comprise receiving a first constituent image of human tissue; receiving a second constituent image of human tissue; overlapping a portion of the second constituent image on a portion of the first constituent image to create an augmented image; and training a model using a dataset comprising at least the augmented image.

Claims (31)

1. A method for improving machine learning algorithm performance comprising:

receiving a first constituent image comprising healthy human tissue;

receiving a second constituent image comprising unhealthy human tissue;

overlapping a portion of the second constituent image on a portion of the first constituent image to create an augmented image;

generating, using a function, a value of a pixel of the augmented image, wherein the pixel is at a grid location in the overlapping portions of the second and the first constituent images within the augmented image, and wherein the function receives as a first input a first value of a pixel of the first constituent image at a grid location within the first constituent image which corresponds to the grid location of the pixel of the augmented image and as a second input a second value of a pixel of the second constituent image at a grid location within the second constituent image which corresponds to the grid location of the pixel of the augmented image; and

training a model using a dataset comprising at least the augmented image.

2. The method of claim 1 , wherein training the model further comprises classifying the augmented image based on a classifier of the first constituent image and on a classifier of the second constituent image.

3. The method of claim 1 , wherein the function may comprise a maximum value of any of the pixels of the constituent images at the location.

4. The method of claim 1 , wherein the function may comprise a weighted sum of the values of the pixels of the constituent images at the location.

5. The method of claim 4 , wherein weights of the weighted sum may be based on at least one of the following: a first image classification, a second image classification, a position in the first constituent image, or a position in the second constituent image.

6. A system for improving machine learning algorithm performance comprising:

a processor;

memory accessible by the processor; and

computer program instructions stored in the memory and executable by the processor to perform the steps of:

receiving a first constituent image comprising healthy human tissue;

receiving a second constituent image comprising unhealthy human tissue:

overlapping a portion of the second constituent image on a portion of the first constituent image to create an augmented image;

generating, using a function, a value of a pixel of the augmented image, wherein the pixel is at a grid location in the overlapping portions of the second and the first constituent images within the augmented image, and wherein the function receives as a first input a first value of a pixel of the first constituent image at a grid location within the first constituent image which corresponds to the grid location of the pixel of the augmented image and as a second input a second value of a pixel of the second constituent image at a grid location within the second constituent image which corresponds to the grid location of the pixel of the augmented image; and

training a model using a dataset comprising at least the augmented image.

7. The system of claim 6 , wherein training the model further comprises classifying the augmented image based on a classifier of the first constituent image and a classifier of the second constituent image.

8. The system of claim 6 , wherein the function may comprise a maximum value of any of the pixels of the constituent images at the location.

9. The system of claim 6 , wherein the function may comprise a weighted sum of the values of the pixels of the constituent images at the location.

10. The system of claim 9 , wherein weights of the weighted sum may be based on at least one of the following: a first image classification, a second image classification, a position in the first constituent image, or a position in the second constituent image.

11. A computer program product for improving machine learning algorithm performance, the computer program product comprising a non-transitory computer readable storage having program instructions embodied therewith, the program instructions executable by a computer, to cause the computer to perform a method comprising:

receiving a first constituent image comprising healthy human tissue;

receiving a second constituent image comprising unhealthy human tissue:

overlapping a portion of the second constituent image on a portion of the first constituent image to create an augmented image;

generating, using a function, a value of a pixel of the augmented image, wherein the pixel is at a grid location in the overlapping portions of the second and the first constituent images within the augmented image, and wherein the function receives as a first input a first value of a pixel of the first constituent image at a grid location within the first constituent image which corresponds to the grid location of the pixel of the augmented image and as a second input a second value of a pixel of the second constituent image at a grid location within the second constituent image which corresponds to the grid location of the pixel of the augmented image; and

training a model using a dataset comprising at least the augmented image.

12. The computer program product of claim 11 , wherein training the model further comprises classifying the augmented image based on a classifier of the first constituent image and a classifier of the second constituent image.

13. The computer program product of claim 11 , wherein the function may comprise a weighted sum of the values of the pixels of the constituent images at the location and wherein the weights of the weighted sum may be based on at least one of the following: a first image classification, a second image classification, a position in the first constituent image, or a position in the second constituent image.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 11, 2021
From: SHOSHAN, YOEL; RATNER, VADIM
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 058079/0530 →
Continuity (1)
Related Publication 20230145270A1 · May 11, 2023
References Cited (17)
US 7558618B1 · Williams · 2009 [cited by examiner]
US 10496902B2 · Inoue · 2019 [cited by examiner]
US 10629305B2 · Muller · 2020 [cited by examiner]
US 10991093B2 · Do · 2021 [cited by examiner]
US 20190087694A1 · Inoue · 2019 [cited by examiner]
US 20200020097A1 · Do · 2020 [cited by examiner]
US 20200117991A1 · Suzuki · 2020 [cited by examiner]
US 20210034921A1 · Pinkovich · 2021 [cited by examiner]
US 20210045838A1 · Bradbury · 2021 [cited by examiner]
US 20210196384A1 · Shelton, IV · 2021 [cited by examiner]
US 20230169666A1 · Pati · 2023 [cited by examiner]
CutMix: Regularization Strategy to Train Strong Classifiers with Localizable Features by Sangdoo Yun, Dongyoon Han, Seong Joon Oh, Sanghyuk Chun, Junsuk Choe, Youngjoon Yoo, published at ICCV 2019, arXiv:1905.04899v2. [cited by applicant]
Seamless Insertion of Pulmonary Nodules in Chest CT Images by Aria Pezeshk, Berkman Sahiner, Rongping Zeng, Adam Wunderlich, Weijie Chen, and Nicholas Petrick, published in IEEE Transactions of Biomedical Engineering De… [cited by applicant]
Seamless lesion insertion for data augmentation in CAD training by Aria Pezeshk, Nicholas Petrick, Weijie Chen, and Berkman Sahiner, published in IEEE Transactions on Medical Imaging. Apr. 2017; 36(4); pp. 1005-1015, do… [cited by applicant]
Seamless lesion insertion in digital mammography: methodology and reader study by Aria Pezeshk, Nicholas Petrick, and Berkman Sahiner published in SPIE Proceedings vol. 9785, Medical Imaging 2016: Computer-Aided Diagnos… [cited by applicant]
Towards the use of computationally inserted lesions for mammographic CAD assessment by Zahra Ghanian, Aria Pezeshk, Nicholas Petrick, and Berkman Sahiner published in SPIE Proceedings vol. 10577, Medical Imaging 2018: I… [cited by applicant]
Zhang et al., Mixup: Beyond Empirical Risk Minimization, Retrieved from: https://arxiv.org/abs/1710.09412, Apr. 27, 2018, vol. 1, 13 pages. [cited by applicant]