IP Library Granted Patent US 12,481,885
Granted Patent B2
US 12,481,885 · App. 18/612,216 · Granted Nov 25, 2025

Systems and methods to train a cell object detector

Inventors: Jack Zeineh (New York, NY); Marcel Prastawa (New York, NY); Gerardo Fernandez (New York, NY)
Assignee: Icahn School of Medicine at Mount Sinai
G06N3/08G06N3/045G06T7/0012G06V10/774G06V10/82G06V20/69G06V20/698G06V30/2504
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,481,885
App. No.
18/612,216
Granted
Nov 25, 2025
Kind
B2
Abstract

Systems and methods to train a cell object detector are described.

Claims (63)

1 . A computer-implemented method comprising:

obtaining training data that includes a first set of cell images and corresponding first labels;

augmenting the training data with a second set of cell images and corresponding second labels;

training a first neural network to perform coarse detection using the augmented training data to detect one or more cell images that include candidate objects; and

training a second neural network to perform fine discrimination to identify an object of a particular type of cell object by providing one or more cell images that include the candidate objects detected by the first neural network as input to the second neural network and labels associated with the one or more cell images, wherein after the training the second neural network can distinguish between objects of a particular type and look-alike objects.

2 . The method of claim 1 , wherein training the second neural network further includes:

receiving an initial outline of a potential object within an image in the first set of cell images;

moving a center of a sample window to each pixel within the initial outline; and

at each pixel within the initial outline:

obtaining a sample of image data corresponding to the sample window;

performing one or more transformations on the sample of image data to generate one or more training images; and

providing the training images to the second neural network as training data.

3 . The method of claim 2 , wherein the transformations are selected from a group of flips, rotations, translations, and combinations thereof.

4 . The method of claim 1 , wherein the corresponding first labels identify one or more types of objects in the first set of cell images including the particular type of cell object and wherein the corresponding second labels identify the one or more look-alike objects in the second set of cell images that are similar in visual appearance to the particular type of cell object.

5 . The method of claim 1 , wherein the second set of cell images includes false positives from a previously trained neural network.

6 . The computer-implemented method of claim 1 , further comprising:

obtaining, as output of the second neural network, an indication of whether the particular type of cell object is present in a cell image.

7 . The method of claim 1 , wherein the first neural network is trained by perturbing original image data using synthetic random transformation of colors or geometry with the corresponding first labels or the corresponding second labels held fixed.

8 . The computer implemented method of claim 1 , wherein the look-alike objects comprise hard negatives.

9 . The computer implemented method of claim 1 , wherein the look-alike objects are stained a similar color shade in the one or more cell images to the color shade of objects of the particular type, but have at least one structural difference from the objects of the particular type.

10 . The computer-implemented method of claim 9 , wherein the at least one structural difference of the look-alike objects is a difference in a texture of a structure.

11 . The computer-implemented method of claim 9 , wherein the at least one structural difference of the look-alike objects is a difference in a color of a structure.

12 . The computer-implemented method of claim 9 , wherein the at least one structural difference of the look-alike objects is a difference in a boundary characteristic.

13 . The computer-implemented method of claim 12 , wherein the at least one structural difference of the look-alike objects is the look-alike objects have a smoother boundary than the objects of the particular type.

14 . The computer-implemented method of claim 1 , wherein the look-alike objects are not distinguishable by the first neural network from the objects of the particular type, but have at least one structural difference from the objects of the particular type.

15 . The computer implemented method of claim 1 , wherein the look-alike objects are stained a similar color shade in the cell images to the color shade of objects of the particular type, but have at least one feature difference from the objects of the particular type.

16 . The computer implemented method of claim 1 , wherein the objects of the particular type comprise mitotic or epi-stroma figures.

17 . A non-transitory computer-readable medium with instructions stored thereon that, when executed by one or more computers, cause the one or more computers to perform operations comprising:

obtaining training data that includes a first set of cell images and corresponding first labels;

augmenting the training data with a second set of cell images and corresponding second labels;

training a first neural network to perform coarse detection using the augmented training data to detect one or more cell images that include candidate objects; and

training a second neural network to perform fine discrimination to identify an object of a particular type of cell object by providing one or more cell images that include the candidate objects detected by the first neural network as input to the second neural network and labels associated with the one or more cell images, wherein after the training the second neural network can distinguish between objects of a particular type and look-alike objects.

18 . The non-transitory computer-readable medium of claim 17 , wherein training the second neural network further includes:

receiving an initial outline of a potential object within an image in the first set of cell images;

moving a center of a sample window to each pixel within the initial outline;

at each pixel location within the initial outline:

obtaining a sample of image data corresponding to the sample window;

performing one or more transformations on the sample of image data to generate one or more training images; and

providing the training images to the second neural network as training data.

19 . The non-transitory computer-readable medium of claim 18 , wherein the transformations are selected from a group of flips, rotations, translations, and combinations thereof.

20 . The non-transitory computer-readable medium of claim 17 , wherein the corresponding first labels identify one or more types of objects in the first set of cell images including the particular type of cell object and wherein the corresponding second labels identify the one or more look-alike objects in the second set of cell images that are similar in visual appearance to the particular type of cell object.

21 . The non-transitory computer-readable medium of claim 17 , wherein the second set of cell images includes false positives from a previously trained neural network.

22 . The non-transitory computer-readable medium of claim 17 , wherein the operations further include:

obtaining, as output of the second neural network, an indication of whether the particular type of cell object is present in a cell image.

23 . The non-transitory computer-readable medium of claim 17 , wherein the first neural network is trained by perturbing original image data using synthetic random transformation of colors or geometry with the corresponding first labels or the corresponding second labels held fixed.

24 . A system, comprising:

one or more processors coupled to a computer readable memory having stored thereon software instructions that, when executed by the one or more processors, cause the one or more processors to perform or control performance of operations including:

obtaining training data that includes a first set of cell images and corresponding first labels;

augmenting the training data with a second set of cell images and corresponding second labels;

training a first neural network to perform coarse detection using the augmented training data to detect one or more cell images that include candidate objects; and

training a second neural network to perform fine discrimination to identify an object of a particular type of cell object by providing one or more cell images that include the candidate objects detected by the first neural network as input to the second neural network and labels associated with the one or more cell images, wherein after the training the second neural network can distinguish between objects of a particular type and look-alike objects.

25 . The system of claim 24 , wherein training the second neural network further includes:

receiving an initial outline of a potential object within an image in the first set of cell images;

moving a center of a sample window to each pixel within the initial outline; and

at each pixel within the initial outline:

obtaining a sample of image data corresponding to the sample window;

performing one or more transformations on the sample of image data to generate one or more training images; and

providing the training images to the second neural network as training data.

26 . The system of claim 25 , wherein the transformations are selected from a group of flips, rotations, translations, and combinations thereof.

27 . The system of claim 24 , wherein the corresponding first labels identify one or more types of objects in the first set of cell images including the particular type of cell object and wherein the corresponding second labels identify the one or more look-alike objects in the second set of cell images that are similar in visual appearance to the particular type of cell object.

28 . The system of claim 24 , wherein the second set of cell images includes false positives from a previously trained neural network.

29 . The system of claim 24 , wherein the operations further include:

obtaining, as output of the second neural network, an indication of whether the particular type of cell object is present in a cell image.

Continuity (3)
Continuation 17613433
Provisional Application 62852184 · May 23, 2019
Related Publication 20240232627A1 · Jul 11, 2024
References Cited (24)
US 4523278A · Reinhardt · 1985 [cited by applicant]
US 5719024A · Cabib · 1998 [cited by applicant]
US 10943346B2 · Wirch · 2021 [cited by applicant]
US 11093729B2 · Ohsaka · 2021 [cited by applicant]
US 11379516B2 · Peng · 2022 [cited by examiner]
US 11966842B2 · Zeineh · 2024 [cited by examiner]
US 20110206262A1 · Sammak · 2011 [cited by applicant]
US 20160110584A1 · Remiszewski · 2016 [cited by applicant]
US 20190057769A1 · Bernard · 2019 [cited by applicant]
US 20190065817A1 · Mesmakhosroshahi et al. · 2019 [cited by applicant]
US 20190258878A1 · Koivisto · 2019 [cited by examiner]
US 20190347467A1 · Ohsaka · 2019 [cited by applicant]
US 20210019342A1 · Peng · 2021 [cited by examiner]
US 20210117729A1 · Bharti · 2021 [cited by applicant]
US 20220215253A1 · Zeineh · 2022 [cited by examiner]
US 20240232627A1 · Zeineh · 2024 [cited by examiner]
CN 105913076A · 2016 [cited by examiner]
WO WO2017075294A1 · 2017 [cited by examiner]
Extended European Search Report In European Application No. 20809263.5, Apr. 25, 2023, 11 pages. [cited by applicant]
“International Preliminary Report of Patentability in PCT/US2020/034320”, Dec. 2, 2021. [cited by applicant]
“International Search Report and Written Opinion in PCT/US2020/034320”, Jul. 31, 2020, 8 Pages. [cited by applicant]
Klauschen F., et al., “Scoring of 1-15 tumor-infiltrating lymphocytes: From visual estimation to machine learning”, Seminars in Cancer Biology., vol. 52, Oct. 1, 2018, pp. 151-157. [cited by applicant]
Wahab Noorul, et al., “Two-phase deep 1-15 G06V convolutional neural network for reducing class skewness in histopathological images based breast cancer detection”, Computers in Biology and Medicine, New York, NY, US, v… [cited by applicant]
Wang Justin L, et al., “Classification of 1-15 White Blood Cells with PatternNet-fused Ensemble of Convolutional Neural Networks (PECNN)”, 2018 IEEE International Symposium on Signal Processing and Information Technolog… [cited by applicant]