IP Library › Granted Patent US 12,272,161
Granted Patent B2
US 12,272,161 · App. 17/596,279 · Granted Apr 8, 2025

System and method for processing biology-related data and a microscope

Inventor: Constantin Kappel (Schriesheim, DE)
Assignee: Leica Microsystems CMS GmbH
G06V20/69G06V10/454G06V10/774G06V10/82G06V20/698
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,272,161
App. No.
17/596,279
Granted
Apr 8, 2025
Kind
B2
Abstract

A system ( 100 ) comprising one or more processors ( 110 ) and one or more storage devices ( 120 ) is configured to obtain biology-related image-based input data ( 107 ) and generate a high-dimensional representation of the biology-related image-based input data ( 107 ) by a trained visual recognition machine-learning algorithm executed by the one or more processors ( 110 ). The high-dimensional representation comprises at least 3 entries each having a different value. Further, the system is configured to at least one of store the high-dimensional representation of the biology-related image-based input data ( 107 ) together with the biology-related image-based input data ( 107 ) by the one or more storage devices ( 120 ) or output biology-related language-based output data ( 109 ) corresponding to the high-dimensional representation.

Claims (48)

1. A system comprising one or more processors, one or more interfaces, and one or more storage devices, wherein the system, using the one or more processors, is configured to:

obtain biology-related image-based input data via the one or more interfaces;

select a trained visual recognition machine-learning algorithm from a plurality of trained visual recognition machine-learning algorithms based on the biology-related image-based input data, wherein the trained visual recognition machine-learning algorithm is selected so that the values of more than 5 entries of the high-dimensional representation are larger than 10% of a largest absolute value of the entries of the high-dimensional representation;

generate a high-dimensional representation of the biology-related image-based input data by the trained visual recognition machine-learning algorithm executed by the one or more processors, wherein the high-dimensional representation comprises at least 3 entries each having a different value; and

at least one of:

store the high-dimensional representation of the biology-related image-based input data together with the biology-related image-based input data by the one or more storage devices; or

output biology-related language-based output data corresponding to the high-dimensional representation.

2. The system of claim 1 , wherein the biology-related image-based input data is image data of an image of at least one of a biological structure comprising a nucleotide sequence, a biological structure comprising a protein sequence, a biological molecule, biological tissue, a biological structure with a specific behavior, or a biological structure with a specific biological function or a specific biological activity.

3. The system of claim 1 , wherein the values of one or more entries of the high-dimensional representation are proportional to a likelihood of a presence of a specific biological function or a specific biological activity.

4. The system of claim 1 , wherein the biology-related language- based output data is at least one of a nucleotide sequence, a protein sequence, a description of a biological molecule or biological structure, a description of a behavior of a biological molecule or biological structure, or a description of a biological function or a biological activity.

5. The system of claim 1 , further comprising a microscope configured to obtain the biology-related image-based input data by taking an image of a biological specimen.

6. The system of claim 1 , wherein the high-dimensional representation is a numerical representation.

7. The system of claim 1 , wherein the high-dimensional representation comprises more than 100 dimensions.

8. The system of claim 1 , wherein the high-dimensional representation is a vector.

9. The system of claim 1 , wherein the trained visual recognition machine-learning algorithm is selected so that more than 50% of values of the entries of the high-dimensional representation are unequal 0.

10. The system of claim 1 , wherein the trained visual recognition machine-learning algorithm comprises a trained visual recognition neural network.

11. The system of claim 10 , wherein the trained visual recognition neural network comprises more than 30 layers.

12. The system of claim 10 , wherein the trained visual recognition neural network is a convolutional neural network or a capsule network.

13. The system of claim 10 , wherein the trained visual recognition neural network comprises a plurality of convolution layers and a plurality of pooling layers.

14. The system of claim 10 , wherein the trained visual recognition neural network uses a rectified linear unit activation function.

15. The system of claim 1 , wherein the system is further configured to determine, using the one or more processors, the biology-related language-based output data based on the high-dimensional representation by a decoder of a trained language recognition machine-learning algorithm executed by the one or more processors.

16. The system of claim 15 , wherein the biology-related language-based output data is an entry of a vocabulary trained by the trained language recognition machine-learning algorithm.

17. The system of claim 1 , wherein the system, using the one or more processors further configured to:

obtain a plurality of biology-related image-based data sets via the one or more interfaces, wherein the plurality of biology-related image-based data sets are stored in a database;

generate an individual high-dimensional representation for every biology-related image-based input data set of the plurality of biology-related image-based data sets by the trained visual recognition machine-learning algorithm executed by the one or more processors; and

at least one of:

store the individual high-dimensional representations together with the corresponding biology-related image-based input data sets by the one or more storage devices, or

output biology-related language-based output data sets corresponding to the individual high-dimensional representations.

18. The system of claim 17 , wherein the system, using the one or more processors, is further configured to:

receive biology-related language-based search data via the one or more interfaces;

generate a high-dimensional representation of the biology-related language-based search data by a trained language recognition machine-learning algorithm executed by the one or more processors;

compare the high-dimensional representation of the biology-related language-based search data with the individual high-dimensional representations of the plurality of biology-related image-based data sets; and

output a biology-related image-based data set of the plurality of biology-related image-based data sets based on the comparison.

19. The system of one of claim 1 , wherein the system, using the one or more processors, is further configured to select the trained visual recognition machine-learning algorithm using a classification algorithm configured to classify the biology-related image-based input data.

20. The system of claim 1 , wherein the system, using the one or more processors configured to:

select a second trained visual recognition machine-learning algorithm from the plurality of trained visual recognition machine-learning algorithms;

generate a second high-dimensional representation of the biology-related image-based input data by the second trained visual recognition machine-learning algorithm executed by the one or more processors, wherein the second high-dimensional representation comprises at least 3 entries each having a different value;

at least one of:

store the second high-dimensional representation of the biology-related image-based input data together with the high-dimensional representation and the biology-related image-based input data by the one or more storage devices, or output the biology-related language-based output data and second biology-related language-based output data corresponding to the second high-dimensional representation.

21. The system of claim 1 , wherein the system is configured to control an operation of a microscope.

22. A microscope comprising a system of claim 1 .

23. A method for processing biology-related image-based input data, the method comprising:

obtaining biology-related image-based input data;

selecting a trained visual recognition machine-learning algorithm from a plurality of trained visual recognition machine-learning algorithms based on the biology-related image-based input data, wherein the trained visual recognition machine-learning algorithm is selected so that the values of more than 5 entries of the high-dimensional representation are larger than 10% of a largest absolute value of the entries of the high-dimensional representation;

generating a high-dimensional representation of the biology-related image-based input data by the trained visual recognition machine-learning algorithm, wherein the high-dimensional representation comprises at least 3 entries each having a different value; and

at least one of:

storing the high-dimensional representation of the biology-related image-based input data together with the biology-related image-based input data, or outputting biology-related language-based output data corresponding to the high-dimensional representation.

24. A non-transitory, computer readable medium having a program code for performing a method according to claim 23 when the program is executed by processor.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 4, 2024
From: KAPPEL, CONSTANTIN
To: LEICA MICROSYSTEMS CMS GMBH
Reel/Frame 069476/0177 →
Continuity (1)
Related Publication 20220254177A1 · Aug 11, 2022
References Cited (31)
US 6789069B1 · Barnhill · 2004 [cited by examiner]
US 10049450B2 · Madabhushi · 2018 [cited by examiner]
US 10127475B1 · Corrado · 2018 [cited by examiner]
US 10146914B1 · Victors · 2018 [cited by examiner]
US 10496884B1 · Nguyen · 2019 [cited by examiner]
US 10740896B2 · Georgescu · 2020 [cited by examiner]
US 10769501B1 · Ando · 2020 [cited by examiner]
US 11017896B2 · Madabhushi · 2021 [cited by examiner]
US 11087864B2 · Xie · 2021 [cited by examiner]
US 11301961B2 · Takeshima · 2022 [cited by examiner]
US 11636399B2 · Kajihara · 2023 [cited by examiner]
US 11901077B2 · Klaiman · 2024 [cited by examiner]
US 20020111742A1 · Rocke · 2002 [cited by examiner]
US 20170249548A1 · Nelson et al. · 2017 [cited by applicant]
US 20180232400A1 · Saavedra Rondo · 2018 [cited by applicant]
US 20180232883A1 · Sethi · 2018 [cited by examiner]
US 20190065817A1 · Mesmakhosroshahi · 2019 [cited by examiner]
US 20190087532A1 · Madabhushi · 2019 [cited by examiner]
US 20190183429A1 · Sung · 2019 [cited by examiner]
US 20190251688A1 · Madabhushi · 2019 [cited by examiner]
US 20240054384A1 · Catron · 2024 [cited by examiner]
CN 102663924A · 2012 [cited by applicant]
CN 109036571A · 2018 [cited by applicant]
JP 2013201909A · 2013 [cited by applicant]
JP 2016517115A · 2016 [cited by applicant]
JP 2018125019A · 2018 [cited by applicant]
WO 2018091486A1 · 2018 [cited by applicant]
WO WO2019199392A1 · 2019 [cited by examiner]
Caleb Vununu et al. A Deep Feature Learning Scheme for Counting the Cells in Microscopy Data. [cited by applicant]
Muramoto et al., Image Caption Generation Using Hint-Words, 2019, Journal of the Information Knowledge Society of Japan, 2019, vol. 29, No. 2, pp. 153-158. [cited by applicant]
Mendieta, Milton et al., “A Cross-Modal transfer approach for histological images: A case study in aquaculture for disease identification using Zero-Shot learning”, IEEE Second Ecuador Technical Chapters Meeting(ETCM), … [cited by applicant]
Cited By (2)
US 12,511,921 US 12,731,661