IP Library › Granted Patent US 12,340,506
Granted Patent B2
US 12,340,506 · App. 17/608,016 · Granted Jun 24, 2025

System and method for attention-based classification of high-resolution microscopy images

Inventors: Saeed Hassanpour (Hanover, NH); Naofumi Tomita (Hanover, NH)
Assignee: The Trustees of Dartmouth College
G06T7/0012G06T7/11G06V10/82G06T2207/10056G06T2207/30004
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,340,506
App. No.
17/608,016
Granted
Jun 24, 2025
Kind
B2
Abstract

This invention provides a system and method for analyzing and classifying imaged from whole slides of tissue. A source of image data transmits images of the tissue on the whole slides to a GPU. The GPU performs a feature extraction process that identifies and segments regions of interests in each of the images, and an attention network that, based upon training from an expert, identifies trained characteristics can comprise cancerous and/or pre-cancerous conditions/e.g. those associated with a gastrointestinal tract, such ad Barret's Esophagus. The feature extraction process can include a convolutional neural network (CNN). The attention network can be adapted performs attention/based weighting of features relative to the trained characteristics, and/or the attention network can include 3D convolutional filters. The image data is acquired using an image sensor having approximately 100 Megapixel resolution.

Claims (32)

1. A system for analyzing and classifying images from whole slides of tissue comprising:

a source of image data including images of the tissue on the whole slides, each of the images divided into a plurality of adjacent non-overlapping tiles;

a computer processor executing a CNN-based feature extraction process by which the processor identifies regions of interest in each non-overlapping tile of the images thereby producing a grid-based feature map of the identified regions of interest; and

an attention network that, based upon training from an expert, when executed by the processor, configures the processor to identify trained characteristics in the regions of interest of the grid-based feature map and provide identification data for a least a portion of the identified regions of interest to a user,

wherein the attention network-configured processor performs attention-based weighting of features relative to the trained characteristics, and

wherein the attention network includes 3D convolutional filters of size N×d×d, where N is a depth of a filter kernel and d denotes a height and width of the kernel.

2. The system as set forth in claim 1 wherein the characteristics comprise medical conditions.

3. The system as set forth in claim 2 wherein the medical conditions comprise at least one of cancerous and pre-cancerous conditions.

4. The system as set forth in claim 3 wherein the tissue is associated with a gastrointestinal tract of the patient.

5. The system as set forth in claim 1 wherein the processor includes a GPU that operates the feature extraction process and executes the attention network.

6. The system as set forth in claim 1 wherein the image data of each whole slide is acquired using an image sensor having at least approximately 100 Megapixel resolution.

7. A method for analyzing and classifying images from whole slides of tissue comprising the steps of:

acquiring image data including images of the tissue on the whole slides;

dividing each of the images into a plurality of adjacent non-overlapping tiles;

extracting features with a CNN-configured processor by identifying regions of interest in each non-overlapping tile of the images; and

based upon training from an expert, identifying, with a processor operating an attention network, trained characteristics in one or more of the regions of interest and providing identification data for at least a portion of the regions of interest to a user

wherein the step of the processor operating the attention network comprises performing attention-based weighting of features relative to the trained characteristics, and

wherein the attention network includes 3D convolutional filters of size N×d×d, where N is a depth of a filter kernel and d denotes a height and width of the kernel.

8. The method as set forth in claim 7 wherein the characteristics comprise at least one of visible tissue-related medical conditions, cancerous conditions and pre-cancerous conditions.

9. The method as set forth in claim 8 wherein the tissue is associated with a gastrointestinal tract of the patient.

10. The method as set forth in claim 7 further comprising a GPU that operates the step of extracting and that operates the attention network.

11. A non-transitory, computer-readable medium including program instructions that when executed by a computer processor perform the steps of:

extracting, with a trained CNN, features from acquired image data, including images of the tissue on whole slides divided into a plurality of adjacent non-overlapping tiles, by identifying regions of interest in each of at least a portion of the tiles of the images; and

based upon training from an expert, identifying, by operating an attention network, trained characteristics in the regions of interest and providing identification data for at least a portion of the regions of interest to a user through an interface,

wherein the step of operating the attention network comprises performing attention-based weighting of features relative to the trained characteristics, and

wherein the attention network applies 3D convolutional filters of size N×d×d to the regions of interest, where N is a depth of a filter kernel and d denotes a height and width of the kernel.

12. The non-transitory, computer-readable medium as set forth in claim 11 wherein the characteristics comprise at least one of visible tissue-related medical conditions, cancerous conditions and pre-cancerous conditions.

13. The non-transitory, computer-readable medium as set forth in claim 12 wherein the tissue is associated with a gastrointestinal tract of the patient.

14. The method as set forth in claim 7 wherein the processor includes a GPU that performs the extracting of features and the operating of the attention network.

15. The method as set forth in claim 7 wherein the image data of each whole slide is acquired using an image sensor having at least approximately 100 Megapixel resolution.

16. The non-transitory, computer-readable medium as set forth in claim 11 wherein the processor includes a GPU that performs the extracting of features and the operating of the attention network.

17. The non-transitory, computer-readable medium as set forth in claim 11 wherein the image data of each whole slide is acquired using an image sensor having at least approximately 100 Megapixel resolution.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 15, 2021
From: HASSANPOUR, SAEED; TOMITA, NAOFUMI
To: THE TRUSTEES OF DARTMOUTH COLLEGE
Reel/Frame 058115/0174 →
Continuity (2)
Provisional Application 62840538 · Apr 30, 2019
Related Publication 20220309653A1 · Sep 29, 2022
References Cited (34)
US 10013781B1 · Gammage · 2018 [cited by examiner]
US 20090069477A1 · Vogt · 2009 [cited by examiner]
US 20150347833A1 · Robinson · 2015 [cited by examiner]
US 20170200066A1 · Wang · 2017 [cited by examiner]
US 20180165934A1 · Pan · 2018 [cited by applicant]
US 20180204048A1 · Chefd'Hotel · 2018 [cited by applicant]
US 20190114770A1 · Song · 2019 [cited by applicant]
US 20190213779A1 · Sutton · 2019 [cited by examiner]
US 20190355113A1 · Wirch · 2019 [cited by examiner]
US 20200160510A1 · Lindemer · 2020 [cited by examiner]
US 20210056722A1 · Wu · 2021 [cited by examiner]
US 20210059762A1 · Ng · 2021 [cited by examiner]
CN 109145927A · 2019 [cited by examiner]
A. Paszke, S. Gross, S. Chintala, and G. Chanan, PyTorch: An Imperative Style, High-Performance Deep Learning Library, ed, 12 pages, 2017. [cited by applicant]
B. Korbar et al., Looking Under the Hood: Deep Neural Network Visualization to Interpret Whole-Slide Image Analysis Outcomes for Colorectal Polyps, in Computer Vision and Pattern Recognition Workshops (CVPRW), 2017 IEEE… [cited by applicant]
Cosatto et al., Automated gastric cancer diagnosis on h&e-stained sections; training a classifier on a large scale with multiple instance machine learning, in Medical Imaging 2013: Digital Pathology, 2013, vol. 8676, p.… [cited by applicant]
D. Komura, S. J. C. Ishikawa, and S. B. Journal, Machine learning methods for histopathological image analysis, vol. 16, pp. 34-42, 2018. [cited by applicant]
E. Pesce, P.P. Ypsilantis, S. Withey, R. Bakewell, V. Goh, and G. J. a. p. a. Montana, Learning to detect chest radiographs containing lung nodules using visual attention networks, 13 pages, 2017. [cited by applicant]
F. Wang et al., Residual attention network for image classification, IEEE, 9 pages, 2017. [cited by applicant]
G. Corredor, J. Whitney, V. L. A. Pedroza, A. Madabhushi, and E. R. J. J. o. M. I. Castro, Training a cell-level classifier for detecting basal-cell carcinoma by combining human visual attention maps with low-level hand… [cited by applicant]
J. Fu, H. Zheng, and T. Mei, Look closer to see better: Recurrent attention convolutional neural network for fine-grained image recognition, in CVPR, 2017, vol. 2, p. 3. [cited by applicant]
K. He, X. Zhang, S. Ren, and J. Sun, Delving deep into rectifiers: Surpassing human-level performance on imagenet classification, in Proceedings of the IEEE international conference on computer vision, 2015, pp. 1026-10… [cited by applicant]
K. He, X. Zhang, S. Ren, and J. Sun, Identity mappings in deep residual networks, in European conference on computer vision, 2016, pp. 630-645: Springer. [cited by applicant]
L. C. Chen, Y. Yang, J. Wang, W. Xu, and A. L. Yuille, Attention to scale: Scale-aware semantic image segmentation, in Proceedings of the IEEE conference on computer vision and pattern recognition, 2016, pp. 3640-3649. [cited by applicant]
L. Hou, D. Samaras, T. M. Kurc, Y. Gao, J. E. Davis, and J. H. Saltz, Patch-based convolutional neural network for whole slide tissue image classification, in Proceedings of the IEEE Conference on Computer Vision and Pa… [cited by applicant]
M. Jaderberg, K. Simonyan, and A. Zisserman, Spatial transformer networks, in Advances in neural information processing systems, Google DeepMind, 2015, pp. 2017-2025. [cited by applicant]
M. Saha, C. Chakraborty, D. J. C. M. I. Racoceanu, and Graphics, Efficient deep learning model for mitosis detection using breast histopathology images, vol. 64, pp. 29-40, 2018. [cited by applicant]
N. Coudray et al., Classification and mutation prediction from non-small cell lung cancer histopathology images using deep learning, vol. 24, No. 10, p. 1559, 2018. [cited by applicant]
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. J. T. J. o. M. L. R. Salakhutdinov, Dropout: a simple way to prevent neural networks from overfitting, vol. 15, No. 1, pp. 1929-1958, 2014. [cited by applicant]
N. Tomita, et al., Finding a Needle in the Haystack: Attention-Based Classification of High Resolution Microscopy Images, 2018, 9 pages. [cited by applicant]
P.P. Ypsilantis and G. J. a. p. a. Montana, Learning what to look in chest X-rays with a recurrent visual attention model, 5 pages, 2017. [cited by applicant]
Q. Guan, Y. Huang, Z. Zhong, Z. Zheng, L. Zheng, and Y. J. a. p. a. Yang, Diagnose like a radiologist: Attention guided convolutional neural network for thorax disease classification, 10 pages, 2018. [cited by applicant]
X. Glorot and Y. Bengio, Understanding the difficulty of training deep feedforward neural networks, in Proceedings of the thirteenth international conference on artificial intelligence and statistics, 2010, pp. 249-256. [cited by applicant]
Y. A. Chung and W. H. J. a. p. a. Weng, Learning Deep Representations of Medical Images using Siamese CNNs with Application to Content-Based Image Retrieval, 8 pages, 2017. [cited by applicant]