EP3510917A1 - Machine learning guided imaging system - Google Patents
Machine learning guided imaging system Download PDFInfo
- Publication number
- EP3510917A1 EP3510917A1 EP18248134.1A EP18248134A EP3510917A1 EP 3510917 A1 EP3510917 A1 EP 3510917A1 EP 18248134 A EP18248134 A EP 18248134A EP 3510917 A1 EP3510917 A1 EP 3510917A1
- Authority
- EP
- European Patent Office
- Prior art keywords
- image
- oct
- machine learning
- horizontal
- imaging
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Pending
Links
Images
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V10/00—Arrangements for image or video recognition or understanding
- G06V10/70—Arrangements for image or video recognition or understanding using pattern recognition or machine learning
- G06V10/82—Arrangements for image or video recognition or understanding using pattern recognition or machine learning using neural networks
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61B—DIAGNOSIS; SURGERY; IDENTIFICATION
- A61B3/00—Apparatus for testing the eyes; Instruments for examining the eyes
- A61B3/10—Objective types, i.e. instruments for examining the eyes independent of the patients' perceptions or reactions
- A61B3/102—Objective types, i.e. instruments for examining the eyes independent of the patients' perceptions or reactions for optical coherence tomography [OCT]
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61B—DIAGNOSIS; SURGERY; IDENTIFICATION
- A61B3/00—Apparatus for testing the eyes; Instruments for examining the eyes
- A61B3/10—Objective types, i.e. instruments for examining the eyes independent of the patients' perceptions or reactions
- A61B3/12—Objective types, i.e. instruments for examining the eyes independent of the patients' perceptions or reactions for looking at the eye fundus, e.g. ophthalmoscopes
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F18/00—Pattern recognition
- G06F18/20—Analysing
- G06F18/24—Classification techniques
- G06F18/241—Classification techniques relating to the classification model, e.g. parametric or non-parametric approaches
- G06F18/2413—Classification techniques relating to the classification model, e.g. parametric or non-parametric approaches based on distances to training or reference patterns
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T12/00—Tomographic reconstruction from projections
- G06T12/30—Image post-processing, e.g. metal artefact correction
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T7/00—Image analysis
- G06T7/0002—Inspection of images, e.g. flaw detection
- G06T7/0012—Biomedical image inspection
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T7/00—Image analysis
- G06T7/0002—Inspection of images, e.g. flaw detection
- G06T7/0012—Biomedical image inspection
- G06T7/0014—Biomedical image inspection using an image reference approach
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T7/00—Image analysis
- G06T7/70—Determining position or orientation of objects or cameras
- G06T7/73—Determining position or orientation of objects or cameras using feature-based methods
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V10/00—Arrangements for image or video recognition or understanding
- G06V10/40—Extraction of image or video features
- G06V10/44—Local feature extraction by analysis of parts of the pattern, e.g. by detecting edges, contours, loops, corners, strokes or intersections; Connectivity analysis, e.g. of connected components
- G06V10/443—Local feature extraction by analysis of parts of the pattern, e.g. by detecting edges, contours, loops, corners, strokes or intersections; Connectivity analysis, e.g. of connected components by matching or filtering
- G06V10/449—Biologically inspired filters, e.g. difference of Gaussians [DoG] or Gabor filters
- G06V10/451—Biologically inspired filters, e.g. difference of Gaussians [DoG] or Gabor filters with interaction between the filter responses, e.g. cortical complex cells
- G06V10/454—Integrating the filters into a hierarchical structure, e.g. convolutional neural networks [CNN]
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V10/00—Arrangements for image or video recognition or understanding
- G06V10/70—Arrangements for image or video recognition or understanding using pattern recognition or machine learning
- G06V10/764—Arrangements for image or video recognition or understanding using pattern recognition or machine learning using classification, e.g. of video objects
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T2207/00—Indexing scheme for image analysis or image enhancement
- G06T2207/10—Image acquisition modality
- G06T2207/10072—Tomographic images
- G06T2207/10101—Optical tomography; Optical coherence tomography [OCT]
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T2207/00—Indexing scheme for image analysis or image enhancement
- G06T2207/20—Special algorithmic details
- G06T2207/20081—Training; Learning
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T2207/00—Indexing scheme for image analysis or image enhancement
- G06T2207/20—Special algorithmic details
- G06T2207/20084—Artificial neural networks [ANN]
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T2207/00—Indexing scheme for image analysis or image enhancement
- G06T2207/30—Subject of image; Context of image processing
- G06T2207/30004—Biomedical image processing
- G06T2207/30041—Eye; Retina; Ophthalmic
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V2201/00—Indexing scheme relating to image or video recognition or understanding
- G06V2201/03—Recognition of patterns in medical or anatomical images
Definitions
- OCT optical coherence tomography
- OCT systems are typically limited to ophthalmologists who can afford the systems and are trained to identify and manually selected region of interest (ROI) for performing OCT imaging.
- ROIs can be identified in advance by knowledgeable specialists (such as ophthalmologists) based on an en-face ophthalmoscopy image (e.g., from fundus imaging).
- fundus imaging may be first used to identify retinal lesions (or other abnormalities) visible on the outer surface of the eye. Regions including these lesions could then be identified as ROIs by the specialist, so that the ROIs can then be subjected to further imaging via OCT.
- the present disclosure relates to a multimodal imaging system and method is capable of taking fundus images, automatically identifying regions of interest (ROIs) of the eye from the fundus images, and performing imaging in the identified ROIs, where the images can provide clinically relevant information for screening purposes.
- ROIs regions of interest
- expert intervention is not necessarily required to perform specialized imaging and thus, such imaging and analysis can be provided at more facilities and for more subjects for a lower cost.
- an imaging method comprises: generating a horizontal image of an object; automatically identifying a region of interest (ROI) of the object based on the horizontal image with a non-fully-supervised machine learning system; and generating a second image of the object within the identified ROI, wherein the second image comprises depth information of the object.
- ROI region of interest
- the horizontal image is a color fundus image; an infrared fundus image; a scanning laser ophthalmoscope (SLO) image; or is derived from 3D optical coherence tomography (OCT) scan data;
- the second image is an OCT image;
- the horizontal image is derived from 3D optical coherence tomography (OCT) scan data, and the second image is an OCT image generated by extracting a portion of the 3D OCT scan data corresponding to the identified ROI;
- the method further comprises discarding portions of the 3D OCT scan data that do not correspond to the identified ROI;
- the method further comprises displaying the second image;
- the method further comprise determining probabilities of identified regions of interest, the probabilities indicating a likelihood that the region of interest represents an abnormality of the object as determined by the non-fully-supervised machine learning system;
- the method further comprises displaying a heat map of the probabilities overlaid on the horizontal image;
- the second image is generated from a plurality of B-scans of
- a method of image analysis with a trained non-fully-supervised machine learning system comprises: receiving a horizontal image of an object from a subject; identifying an abnormality of the object as an output of the trained non-fully-supervised machine learning system based on the received horizontal image; extracting information of the trained non-fully-supervised machine learning used to identify the abnormality; identifying a region of interest within the horizontal image as a region of the horizontal image that contributed to the identification of the abnormality, wherein the non-fully-supervised machine learning system is trained with a plurality of horizontal images of the object from different subjects to identify the abnormality of the object.
- the trained non-fully-supervised machine learning system is a convolutional neural network; the abnormality is a retinopathy disorder; the information of the trained non-fully-supervised machine learning system is extracted by determining class activation maps; and/or the region of interest is identified by comparing pixel values of the determined class activation maps to a predetermined threshold.
- the present description is generally directed to multimodal imaging systems and methods capable of taking fundus images, automatically identifying regions of interest (ROIs) of the eye from the fundus images, and performing OCT imaging in the identified ROIs that can provide clinically relevant information for screening purposes.
- ROIs regions of interest
- the system and methods described herein may provide automated color fundus plus OCT imaging that does not require expert intervention.
- imaging and analysis can be provided at more facilities and for more subjects.
- the description is not limited to fundus and OCT imaging, or even to ophthalmological imaging. Rather, the features described herein could be applied to any complementary imaging modalities, or methods with a common modality, such as MRI, CT, ultrasound, and the like; and to any physiological structures or other objects.
- Automatically identifying ROIs comprises automatically detecting retinal abnormalities in the fundus image.
- the scope of abnormalities that can be detected affects the usability of the resulting imaging system. For example, if the system is only capable of detecting one or a few types of lesions, it would provide little help in identifying an unknown retinopathy (or other disease) of a subject unless that subject's retinopathy happens to be one of the few that the system is capable of detecting. On the other hand, if the system is capable of identifying many types of lesions but takes a long time to analyze (e.g., if one simply combined many specific lesion-specific detection processes), it would provide little help where speed, affordability, and ease of use is desired. In other words, the automatic detection of retinal abnormalities and identification of regions of interest described herein takes into consideration both generality and efficiency of the system.
- automatic detection of retinal abnormalities and identification of ROIs is performed with machine learning systems.
- a subject's eye is imaged (e.g., by fundus imaging or the like) and the resulting image is input to the machine learning system.
- the output of the machine learning system provides useful data for further analysis or imaging.
- Some machine learning techniques such as deep learning are able to identify that a subject's eye is not healthy, but are not able to identify the particular retinopathy or particular ROIs of the eye/image.
- This limitation is caused by the fact that, in supervised machine learning, the machine is trained to correctly predict targets (in this case whether the subject is healthy) based on the input (in this case the fundus image of the subject).
- targets in this case whether the subject is healthy
- the machine In order for the machine to predict the location of lesions, one needs to first train it with images labeled at a pixel level. In other words, each pixel in the image is labeled to indicate whether it is part of an imaged lesion. Because this approach is labor intensive and sensitive to the annotator's knowledge, one lesion may be easily missed, which could significantly degrade the sensitivity of the system.
- weakly supervised (or, non-fully-supervised) machine learning may help overcome this problem.
- weakly supervised learning instead of outputting the prediction of the target (whether the subject is healthy), information regarding how the prediction is made is extracted from the learnt system.
- the extracted information can be the location of the lesion or abnormality that the system recognizes and would use to identify a subject as unhealthy.
- Such a non-fully-supervised machine learning technique can thus, for example, automatically identify a region of interest in an input fundus image, and guide imaging with a second modality (e.g. OCT) in that region of interest.
- OCT second modality
- such weakly supervised machine learning systems may provide general purpose retinal abnormality detection, capable of detecting multiple types of retinopathy.
- the systems and methods herein do not depend on the disease type and can be applied to all subjects. This can be particularly helpful for screening purposes.
- a fundus or like image is captured 100.
- the fundus image is then input into a neural network or other machine learning system 102, which analyzes the fundus image and determines, for example, the existence of a particular retinopathy.
- a neural network or other machine learning system 102 which analyzes the fundus image and determines, for example, the existence of a particular retinopathy.
- information extracted from the machine learning system e.g., information relating how the machine determined a particular output based on an input
- one or more regions of interest are identified 104.
- the present disclosure recognizes that how a machine learning system produces an output (e.g., a retinopathy) based on an input image (e.g., a fundus image of the eye) can be used to identify regions of the image (and correspondingly, the eye) having an abnormality that likely caused the machine learning system to output the particular retinopathy associated with the abnormality. Once identified, those regions can then be more closely studied 106 for example, with additional imaging.
- an output e.g., a retinopathy
- the machine learning system may include a neural network.
- the neural network may be of any type, such as a convolutional neural network (CNN), which is described herein as an example but is not intended to be limiting.
- CNN convolutional neural network
- the CNN is trained to distinguish input images (e.g., color fundus images) of healthy and sick eyes. In other words, "under the hood," the CNN constructs models of what fundus images of healthy and sick eyes look like.
- This framework is illustrated by the flowchart in Fig. 2 .
- a deep convolution neural network 200 is trained by inputting known images of healthy eyes 202 and sick eyes 204.
- the neural network 200 is able to construct a model 206 what a healthy eye image looks like and a model 208 of what a sick eye image looks like.
- the sick eye model 208 is able to recognize 210 portions of eye images that match known sick eye images.
- the neural network 200 is able to output 212 a determination of whether any input image matches the healthy eye model 206 or the sick eye model 208, and what portions of the image match the sick eye model 208. From this, regions where there is an abnormality associated with a retinopathy can be identified; for example, as the regions where an input fundus image matches the sick eye model 208.
- CNNs are a type of machine learning modeled after the physiological vison system of humans.
- the core of a CNN comprises convolution layers including a filter and an activation map.
- the filter also known as a kernel
- the patch (having 3 ⁇ 3 pixels) is the same size as the filter.
- Applying the filter to the entire input image generates the activation values for each pixel of the activation map.
- the activation value is determined by performing a convolution operation on the pixel values of the patch of the input image and the filter.
- the filter effectively "filters" the input image to an activation map based on its content.
- the particular combination of filters and convolution layers constitute the particular machine learning model.
- the convolution operation sums the product of the value of each filter pixel value and the value of the corresponding input image pixel and assigns the summation value as the activation value for a pixel of the activation map corresponding to the middle pixel of the filter.
- the operation corresponds to a pixel-wise multiplication between the patch and the filter.
- the CNN may stack multiple convolution layers in series, effectively applying multiple filters, each with a different purpose/function to the input image.
- the filters of the CNN can be designed to be capable of identifying complex objects.
- the CNN can be trained by inputting images (e.g., fundus images) and a known retinopathy (e.g., healthy or sick, including an identification of a particular disease) associated with each image.
- a known retinopathy e.g., healthy or sick, including an identification of a particular disease
- the CNN learns the set of filters that best separates images of healthy and sick subjects and estimates a probability that the subject is sick/has a particular retinopathy.
- information of retinal abnormalities can be found in the learnt filters of a trained CNN for a particular retinopathy.
- learnt filters contain the information that can be used to identify potential regions of interests (e.g., by identifying locations where lesions appear in the fundus image). When this information is extracted from the learnt filters, it can then be applied back to the input image to identify those regions of interest by identifying which portions of the input image match the sick models.
- class activation maps or like methods can be used to retrieve the information in the learnt filters of a CNN.
- the descriptions herein related to CAMs are merely examples, and the present disclosure is not limited to CAMs; rather, any method for extracting information of a learnt neural network or other machine learning algorithm may be used.
- a CAM is retrieved by attaching a global activation pooling (GAP) layer to the final convolution layer of the CNN.
- GAP global activation pooling
- the GAP reduces the final activation maps having many pixels into a single (or at least fewer) representative value(s).
- Fig. 4 illustrates an example of a CNN with multiple convolution layers and an attached GAP layer.
- a fundus image is input to the CNN, which at a first convolutional layer applies a plurality of filters to generate a corresponding plurality of activation maps (three shown). Each of these activation maps is then applied as an input to additional convolution layers.
- a plurality of filters is again applied to generate a corresponding plurality of activation maps (five shown, identified as A 1 - A 4 and A k ) that are used to determine a corresponding plurality of GAPs (for example, according to Equation 1).
- the GAPs are used to determine probabilities of whether the input image is of a healthy or a sick eye.
- w healthy k and w sick k are weights connecting the classification and the different GAPs G k .
- G k indicates the presence of drusen
- w healthy k could be negative and w sick k could be positive.
- the weights may be randomly initialized and adjusted during training of the machine learning system such that z healthy is higher for healthy training images and z sick is higher for diseased training images.
- C i , j can further be rescaled to be ⁇ [0,1].
- Fig. 5 is a heat map indicating detected abnormality regions and probabilities of an abnormality of the CAM as superimposed on an input fundus image.
- the heat maps herein are images of the eye (e.g., the horizontal, surface, or en face images) having a color, shade, hue, or the like corresponding to a probability that the region includes an abnormality. This produces a resultant image of the eye whereby the probability of an abnormality in any particular portion of the eye is represented as a color, shade, hue, or the like in the image.
- Regions 500 each represent a detected abnormal region, and region 510 (having a darker shade) represents a higher probability of abnormality than regions 500.
- the heat maps can be in color such that, for example, color contours highlight the detected abnormal regions and indicate a probability where a color transition from blue to red may indicate an increase in abnormality probability.
- a threshold of 0.4 (where C i,j > 0.4) was used to identify the regions 500.
- Figs. 6-8 further illustrate example CAM heat maps for various retinopathies.
- Fig. 6 is a heat map for an image of an eye having hypertensive and arteriosclerotic retinopathies.
- the CAM overlay indicates regions 600 and 604 as likely abnormal with regions therein 602 and 606 as having the highest probability of an abnormality. Portions of the image in these high-probability regions 602 and 606 are enlarged, where the structural abnormality can be visually confirmed.
- Fig. 7 is a heat map for an image of an eye having micro-aneurysms.
- the CAM map indicated region 700 as being abnormal, and therein, region 702 as a having a high probability of being the location of the abnormality.
- a portion of high-probability region 702 is enlarged 704, which shows visual confirmation of the retinopathy.
- Fig. 8 is a heat map for an image of an eye having a background diabetic retinopathy. Again, the CAM map indicated regions 800 as likely containing abnormalities, with region 802 having the highest probability of having an abnormality. A portion 804 of high-probability region 802 is enlarged and visually confirms that the abnormality exists.
- a second image can be generated in and/or around the identified ROIs.
- the second image can be generated by a scan with a second modality (e.g., an OCT scan) that provides more detailed imaging, analysis, and information about the retinopathy.
- a second modality e.g., an OCT scan
- an OCT scan may provide a 3D imaging volume at high resolutions so that the internal structure of retinal tissue may be analyzed, whereas the initial en-face image only images the surface of the structure.
- the second image can be an OCT-angiography (OCTA) image, visual field test results, fluorescent angiography, or fluorescent angiography fundus image. Examples of the application of CAMs according to the present disclosure are illustrated in Figs. 9-11 .
- a horizontal image of the retina is taken 902 after the imaging modality used to take the image has been automatically positioned and focused 900 on the retina.
- a "horizontal" image means a surface or en-face image of the object being imaged, (e.g., the retina).
- Such an image may be taken with, for example, a fundus camera (color or infrared), scanning laser ophthalmoscope (SLO), or be a surface image derived from a 3D-OCT scan.
- SLO scanning laser ophthalmoscope
- other modalities and techniques may be used for horizontal/surface images, and the above examples are not limiting.
- ROIs are identified from the horizontal images 904 using the above-described non-fully-supervised machine learning and CAMs. Based on these identifications, OCT imaging and measurement is performed 906 on portions of the retina corresponding to the identified ROI locations of the horizontal image.
- This second imaging of the ROIs may be performed automatically (e.g., OCT imaging may be automatically controlled upon determination of the ROI) or manually instituted by a user.
- OCT imaging OCT imaging
- measurements and/or the horizontal imaging is finally reported to a user and stored 908 for future analysis or review.
- This and other data derived from the method can also be stored, analyzed, and/or reported, for example, in any form of memory, as part of a database or the like (e.g., for future analysis or normative comparisons).
- the reports may include any of the images, heat maps/CAMs, identification of possible disease/retinopathy, and the like.
- a 3D OCT volume of an eye is initially taken and used to obtain the horizontal image for identifying ROIs.
- a second imaging scan need not be performed because all of the relevant data is captured in the initial 3D OCT volume.
- the OCT imaging modality is initially positioned and focused 1000, and then the 3D OCT volume is acquired 1002.
- a horizontal image is obtained 1004.
- the horizontal image may be obtained by any technique, for example, flattening the volume along a depth dimension by averaging the pixel values across a relevant depth (Z) at a particular X-Y location of the volume.
- ROIs are identified 1004 using machine learning and CAMs.
- the locations of the ROIs are then translated to the original 3D OCT volume 1006 so that the relevant volumetric data corresponding to the ROI can be extracted and/or otherwise highlighted 1008. All of the information including the entire 3D OCT image data can also be stored, analyzed, and/or reported; or, alternatively, the remainder of 3D image data not associated with the ROIs can be otherwise discarded.
- the identified ROIs can also be useful 3D OCT scans are performed subsequently to the ROI identification. This is because the horizontal resolution of 3D OCT volumes is inversely proportional to the scan area. Thus, ROIs can guide future OCT scans at higher resolution in the most relevant regions by limiting the scan area to those most relevant regions. In other words, ROIs derived from an initial survey OCT scan that covers a large area, or similar horizontal images from large area scans from a different imaging modality, can be used to generate higher resolution scans in and around the identified ROIs by limiting the scan area. Additionally, instead of overwhelming users with a large 3D volume of data, B-scans can be selected from the ROIs that highlight an anomaly therein.
- Fig. 11 illustrates a third application in accordance with the above.
- the application of Fig. 11 is similar to that of Fig. 8 , however, a 3D OCT survey image covering a large area of an eye is initially taken 1102 (after automatic positioning and focusing 1100) and used to obtain the horizontal image of the retina 1104.
- denser (or higher resolution) 3D OCT images are taken 1108 of the retina at locations corresponding to the ROIs to form the second images.
- denser images can reveal finer and more granular details of the tissue and can better support a particular diagnosis or disease identification, and better aid in analyzing progression of a disease.
- the dense 3D OCT images, and/or survey image are stored, analyzed, and/or reported 1110.
- an example system corresponding to the disclosure herein is schematically illustrated in Fig. 12 and comprises a first imaging modality 1200 that is capable of generating a horizontal image, a second imaging modality 1202 capable of generating images of regions of interest identified in the horizontal image, and a computer 1204 having a processor 1206, or the like configured to automatically identify the regions of interest in the horizontal image according to the above method.
- computer further includes at least one processor (e.g., a central processing unit CPU, graphics processing unit GPU, or the like) that is capable of machine learning with, for example, the above-described CNN that forms a machine learning system 1212.
- the processor of the machine learning system 1212 may be separate from or integrated with the processor 1206 of the computer 1204.
- the computer could also be configured with an input interface 1210 to receive input images from a user, or directly from the first or second imaging modalities; and an output interface 1210 such as a display to output the images taken, and the data collected to the user, or to directly send the ROI information to the second imaging modality.
- these outputs may be the raw CAM or CNN data, heat maps, and the like.
- the system may also include memory 1208, such as RAM, ROM, flash memory, hard disks, and the like for storing the images and associated data.
- the first and second modalities may be the same (and comprised of common hardware features), for example, if the horizontal image and ROI image data are both (or come from) 3D OCT volume data sets collected from a single scan (as in the embodiment of Fig. 10 ).
- the processor 1206, memory 1208, computer 1204, and/or the like may be integrated with the imaging modalities (or lone modality) 1200, 1202, or wholly separate and simply supplied with imaging data to be analyzed.
- the elements of the computer 1204 may also be fully integrated into a single device, or separated as multiple devices, for example, if the machine learning system 1212 is embodied on a separate computer device.
- Table 1 below illustrates the specificity and sensitivity for various configurations of models and trainable convolution layers using roughly 400 fundus images having a resolution of 500 ⁇ 500 for training, with 39 possible retinopathy manifestations.
- Conv(X) in the table refers to a convolution layer with X number of filters. Table 1.
- Example machine learning configurations Model Configuration # of trainable convolution layers Specificity Sensitivity Xception Add an additional Conv(20) to the end of the model 1 87% 30% Replace the final convolution layer with 1 80% 47% Conv(20) 2 93% 37% Inception (V3) Add an additional Conv(2048) layer to the end of the model 1 81% 34% ResNet50 Add an additional Conv(20) layer after the final activation layer 1 88% 47% All layers 83% 48% Add an additional Conv(20) layer before the final activation layer 1 91% 44% All layers 92% 14% InceptionResNetV2 Add an additional Conv(20) to the end of the model 1 87% 63% ⁇ 20% of all layers 79% 41% MobileNet Add an additional Conv(1024) to the end of the model All layers 83% 31% Replace the final convolution layer with Conv(1024) 1 62% 68% All layers 81% 34% VGG16 Add an additional Conv(512) to the end of the model 1 93% 91% 4 86% 98% 7
- the training models achieved a good sensitivity and specificity.
- the systems and methods disclosed herein are capable of achieving high sensitivity and specificity while utilizing single models for identifying 39 different retinopathy manifestations. This success is possible with a variety of different models. While 39 retinopathy manifestations were tested, a more complex dataset (one with more retinopathy manifestations) could be used to provide the high sensitivity and specificity with more retinopathies. Thus there is no limit to the number of retinopathies to which the present disclosure can be applied.
- the aspects of the present disclosure may be used with other models, machine learning algorithms, and methods of extracting information from those models, including those designed specifically for use with the present disclosure.
- example images shown in Figs. 4-8 were formed from the VGG19 model, where the final convolution layer was replaced with Conv(512), with four trainable convolution layers. This configuration is identified with an asterisk in the above table.
- the present invention may also relate to the following embodiments:
Landscapes
- Engineering & Computer Science (AREA)
- Health & Medical Sciences (AREA)
- Theoretical Computer Science (AREA)
- Physics & Mathematics (AREA)
- Life Sciences & Earth Sciences (AREA)
- General Health & Medical Sciences (AREA)
- General Physics & Mathematics (AREA)
- Computer Vision & Pattern Recognition (AREA)
- Medical Informatics (AREA)
- Evolutionary Computation (AREA)
- Artificial Intelligence (AREA)
- Biomedical Technology (AREA)
- Molecular Biology (AREA)
- Multimedia (AREA)
- Nuclear Medicine, Radiotherapy & Molecular Imaging (AREA)
- Radiology & Medical Imaging (AREA)
- Veterinary Medicine (AREA)
- Public Health (AREA)
- Animal Behavior & Ethology (AREA)
- Surgery (AREA)
- Heart & Thoracic Surgery (AREA)
- Ophthalmology & Optometry (AREA)
- Biophysics (AREA)
- Data Mining & Analysis (AREA)
- Computing Systems (AREA)
- Databases & Information Systems (AREA)
- Software Systems (AREA)
- Quality & Reliability (AREA)
- Biodiversity & Conservation Biology (AREA)
- General Engineering & Computer Science (AREA)
- Evolutionary Biology (AREA)
- Bioinformatics & Computational Biology (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Eye Examination Apparatus (AREA)
- Image Analysis (AREA)
Abstract
Description
- This application claims priority to
, filed on December 28, 2017, entitled "MACHINE LEARNING GUIDED IMAGING SYSTEM", the entirety of which is incorporated herein by reference.U.S. Provisional Application Serial No. 62/611,352 - There exist various modalities for imaging the interior of the eye. The information obtained from these modalities can be used to diagnose the state of health of the eye. If combined, the information derived from these modalities can yield important clues as to the diagnosis and prognosis of disease. For example, fundus imaging is a technique that covers a large field of view in one measurement, but only images the outer surface of the eye. Because fundus imaging lacks depth information, fundus imaging by itself does not enable further assessment on abnormalities to the interior of the eye. On the other hand, optical coherence tomography (OCT), for example, can provide the depth information. However, the field of view of OCT can be limited and thus can require one to specify a particular scanning region. While 3D OCT exists for larger volumes, the data size is often too large to analyze and manage for vision screening purposes.
- Further, due to high costs and technical knowledge often needed to operate OCT systems, OCT systems are typically limited to ophthalmologists who can afford the systems and are trained to identify and manually selected region of interest (ROI) for performing OCT imaging. These ROIs can be identified in advance by knowledgeable specialists (such as ophthalmologists) based on an en-face ophthalmoscopy image (e.g., from fundus imaging). For example, fundus imaging may be first used to identify retinal lesions (or other abnormalities) visible on the outer surface of the eye. Regions including these lesions could then be identified as ROIs by the specialist, so that the ROIs can then be subjected to further imaging via OCT.
- While OCT systems have become more affordable and available for use together with traditional a fundus cameras, many users are still not experienced enough to take full advantage of the capabilities of both imaging modalities. In particular, it is challenging to find an appropriate ROI for the OCT imaging based on an en-face fundus image. This difficulty is exacerbated if imaging is done for screening purposes, where time is limited and the type of disease, if any, is unknown. Due to this, selecting an appropriate ROI is subject to human error and is constrained by the user's knowledge. Even to the extent ROI selection has been automated, the automation is still based on a set of manually defined rules (e.g., colors, orientations, area size), which may only be based on or useful for identifying a particular known disease. Because the manually defined rules are unique to each algorithm and each disease, they are limited in their applicability such that many different analyses have to be performed if the disease is unknown.
- Consequently, to the extent combined fundus/OCT imaging systems have been proposed, they still suffer from forms of the above deficiencies.
- In view of the above, the present disclosure relates to a multimodal imaging system and method is capable of taking fundus images, automatically identifying regions of interest (ROIs) of the eye from the fundus images, and performing imaging in the identified ROIs, where the images can provide clinically relevant information for screening purposes. By automatically identifying the ROIs, expert intervention is not necessarily required to perform specialized imaging and thus, such imaging and analysis can be provided at more facilities and for more subjects for a lower cost.
- According to a first example, an imaging method comprises: generating a horizontal image of an object; automatically identifying a region of interest (ROI) of the object based on the horizontal image with a non-fully-supervised machine learning system; and generating a second image of the object within the identified ROI, wherein the second image comprises depth information of the object. In various embodiments of the above example, the horizontal image is a color fundus image; an infrared fundus image; a scanning laser ophthalmoscope (SLO) image; or is derived from 3D optical coherence tomography (OCT) scan data; the second image is an OCT image; the horizontal image is derived from 3D optical coherence tomography (OCT) scan data, and the second image is an OCT image generated by extracting a portion of the 3D OCT scan data corresponding to the identified ROI; the method further comprises discarding portions of the 3D OCT scan data that do not correspond to the identified ROI; the method further comprises displaying the second image; the method further comprise determining probabilities of identified regions of interest, the probabilities indicating a likelihood that the region of interest represents an abnormality of the object as determined by the non-fully-supervised machine learning system; the method further comprises displaying a heat map of the probabilities overlaid on the horizontal image; the second image is generated from a plurality of B-scans of the region of interest; the horizontal image is derived from a 3D survey image and the second image has a greater density than the horizontal image; the horizontal image is a 3D survey image, and the second image is a 3D optical coherence tomography (OCT) image taken of the identified ROI; the method further comprises only storing data corresponding to the second image, or discarding data corresponding to the horizontal image that is not associated with the identified ROI; the object is an eye; the non-fully-supervised machine learning system comprises a convolutional neural network; and/or the ROI is identified by obtaining a class activation map of the non-fully-supervised machine learning system.
- According to another example, a method of image analysis with a trained non-fully-supervised machine learning system comprises: receiving a horizontal image of an object from a subject; identifying an abnormality of the object as an output of the trained non-fully-supervised machine learning system based on the received horizontal image; extracting information of the trained non-fully-supervised machine learning used to identify the abnormality; identifying a region of interest within the horizontal image as a region of the horizontal image that contributed to the identification of the abnormality, wherein the non-fully-supervised machine learning system is trained with a plurality of horizontal images of the object from different subjects to identify the abnormality of the object.
- According to various embodiments of the second example, the trained non-fully-supervised machine learning system is a convolutional neural network; the abnormality is a retinopathy disorder; the information of the trained non-fully-supervised machine learning system is extracted by determining class activation maps; and/or the region of interest is identified by comparing pixel values of the determined class activation maps to a predetermined threshold.
-
-
Figure 1 illustrates an example operation of the systems and methods described herein. -
Figure 2 illustrates an example convolutional neural network framework. -
Figure 3 illustrates an example convolution layer of a convolutional neural network. -
Figure 4 illustrates an example of a convolutional neural network with multiple convolution layers and an attached global activation pooling layer. -
Figure 5 is an example heat map indicating the detected abnormality regions and probabilities of an abnormality. -
Figure 6 is an example heat map for an image of an eye having hypertensive and arteriosclerotic retinopathies. -
Figure 7 is an example heat map for an image of an eye having micro-aneurysms. -
Figure 8 is an example heat map for an image of an eye having a background diabetic retinopathy. -
Figure 9 is a flow chart of an example application of class activation maps. -
Figure 10 is a flow chart of another example application of class activation maps. -
Figure 11 is a flow chart of another example application of class activation maps. -
Figure 12 is a schematic diagram of an example system described herein. - In view of the above, the present description is generally directed to multimodal imaging systems and methods capable of taking fundus images, automatically identifying regions of interest (ROIs) of the eye from the fundus images, and performing OCT imaging in the identified ROIs that can provide clinically relevant information for screening purposes. In so doing, the system and methods described herein may provide automated color fundus plus OCT imaging that does not require expert intervention. Thus, such imaging and analysis can be provided at more facilities and for more subjects. Of course, the description is not limited to fundus and OCT imaging, or even to ophthalmological imaging. Rather, the features described herein could be applied to any complementary imaging modalities, or methods with a common modality, such as MRI, CT, ultrasound, and the like; and to any physiological structures or other objects.
- Automatically identifying ROIs comprises automatically detecting retinal abnormalities in the fundus image. The scope of abnormalities that can be detected affects the usability of the resulting imaging system. For example, if the system is only capable of detecting one or a few types of lesions, it would provide little help in identifying an unknown retinopathy (or other disease) of a subject unless that subject's retinopathy happens to be one of the few that the system is capable of detecting. On the other hand, if the system is capable of identifying many types of lesions but takes a long time to analyze (e.g., if one simply combined many specific lesion-specific detection processes), it would provide little help where speed, affordability, and ease of use is desired. In other words, the automatic detection of retinal abnormalities and identification of regions of interest described herein takes into consideration both generality and efficiency of the system.
- According to embodiments of the present disclosure, automatic detection of retinal abnormalities and identification of ROIs is performed with machine learning systems. Therewith, a subject's eye is imaged (e.g., by fundus imaging or the like) and the resulting image is input to the machine learning system. The output of the machine learning system provides useful data for further analysis or imaging.
- Some machine learning techniques such as deep learning are able to identify that a subject's eye is not healthy, but are not able to identify the particular retinopathy or particular ROIs of the eye/image. This limitation is caused by the fact that, in supervised machine learning, the machine is trained to correctly predict targets (in this case whether the subject is healthy) based on the input (in this case the fundus image of the subject). In order for the machine to predict the location of lesions, one needs to first train it with images labeled at a pixel level. In other words, each pixel in the image is labeled to indicate whether it is part of an imaged lesion. Because this approach is labor intensive and sensitive to the annotator's knowledge, one lesion may be easily missed, which could significantly degrade the sensitivity of the system.
- By contrast, weakly supervised (or, non-fully-supervised) machine learning as disclosed hereinafter may help overcome this problem. With weakly supervised learning, instead of outputting the prediction of the target (whether the subject is healthy), information regarding how the prediction is made is extracted from the learnt system. For example, the extracted information can be the location of the lesion or abnormality that the system recognizes and would use to identify a subject as unhealthy. Such a non-fully-supervised machine learning technique can thus, for example, automatically identify a region of interest in an input fundus image, and guide imaging with a second modality (e.g. OCT) in that region of interest. In other words, such weakly supervised machine learning systems may provide general purpose retinal abnormality detection, capable of detecting multiple types of retinopathy. As a result, the systems and methods herein do not depend on the disease type and can be applied to all subjects. This can be particularly helpful for screening purposes.
- Briefly, as illustrated in
Fig. 1 , the systems and methods described operate as follows. First, a fundus or like image is captured 100. The fundus image is then input into a neural network or othermachine learning system 102, which analyzes the fundus image and determines, for example, the existence of a particular retinopathy. Using information extracted from the machine learning system (e.g., information relating how the machine determined a particular output based on an input), one or more regions of interest are identified 104. In other words, the present disclosure recognizes that how a machine learning system produces an output (e.g., a retinopathy) based on an input image (e.g., a fundus image of the eye) can be used to identify regions of the image (and correspondingly, the eye) having an abnormality that likely caused the machine learning system to output the particular retinopathy associated with the abnormality. Once identified, those regions can then be more closely studied 106 for example, with additional imaging. - As noted above, the machine learning system may include a neural network. The neural network may be of any type, such as a convolutional neural network (CNN), which is described herein as an example but is not intended to be limiting. The CNN (or other neural network) is trained to distinguish input images (e.g., color fundus images) of healthy and sick eyes. In other words, "under the hood," the CNN constructs models of what fundus images of healthy and sick eyes look like. This framework is illustrated by the flowchart in
Fig. 2 . As seen therein, a deep convolutionneural network 200 is trained by inputting known images ofhealthy eyes 202 andsick eyes 204. Based on these known 202, 204, theimages neural network 200 is able to construct amodel 206 what a healthy eye image looks like and amodel 208 of what a sick eye image looks like. At a high level, thesick eye model 208 is able to recognize 210 portions of eye images that match known sick eye images. Once trained theneural network 200 is able to output 212 a determination of whether any input image matches thehealthy eye model 206 or thesick eye model 208, and what portions of the image match thesick eye model 208. From this, regions where there is an abnormality associated with a retinopathy can be identified; for example, as the regions where an input fundus image matches thesick eye model 208. - CNNs are a type of machine learning modeled after the physiological vison system of humans. As illustrated in
Fig. 3 , the core of a CNN comprises convolution layers including a filter and an activation map. As shown therein, the filter (also known as a kernel) looks at a small patch of an input image (having 6×6 pixels in the example ofFig. 3 ) at a time, and calculates an activation value for a corresponding pixel in the activation map. The patch (having 3×3 pixels) is the same size as the filter. Applying the filter to the entire input image generates the activation values for each pixel of the activation map. The activation value is determined by performing a convolution operation on the pixel values of the patch of the input image and the filter. Thus, the closer the pattern of a small patch of the input image matches the pattern of the filter, the higher the activation value; conversely the less they match, the lower the activation value. Of course, this relationship could be reversed based on the operation, so long as the meaning of the resultant value is understood. In this way, the filter effectively "filters" the input image to an activation map based on its content. The particular combination of filters and convolution layers constitute the particular machine learning model. - According to the example of
Fig. 3 , the convolution operation sums the product of the value of each filter pixel value and the value of the corresponding input image pixel and assigns the summation value as the activation value for a pixel of the activation map corresponding to the middle pixel of the filter. In other words, the operation corresponds to a pixel-wise multiplication between the patch and the filter. Thus, the activation value for the pixel in the second row and second column of the activation map (performed on a patch including the first three rows and columns of the input image identified by the bold outline) in the example ofFig. 3 is equal 0×4 + 0×0 + 0×0 + 0×0 + 1×0 + 1×0 + 0×0 + 1×0 + 2×(-4) = -8. - While the shape of a filter may seem limited and constrained by its size, the CNN may stack multiple convolution layers in series, effectively applying multiple filters, each with a different purpose/function to the input image. Thus the filters of the CNN can be designed to be capable of identifying complex objects.
- As noted above, the CNN can be trained by inputting images (e.g., fundus images) and a known retinopathy (e.g., healthy or sick, including an identification of a particular disease) associated with each image. During training, the CNN learns the set of filters that best separates images of healthy and sick subjects and estimates a probability that the subject is sick/has a particular retinopathy. Thus, information of retinal abnormalities can be found in the learnt filters of a trained CNN for a particular retinopathy. In other words, learnt filters contain the information that can be used to identify potential regions of interests (e.g., by identifying locations where lesions appear in the fundus image). When this information is extracted from the learnt filters, it can then be applied back to the input image to identify those regions of interest by identifying which portions of the input image match the sick models.
- To this end, class activation maps (CAMs) or like methods can be used to retrieve the information in the learnt filters of a CNN. The descriptions herein related to CAMs are merely examples, and the present disclosure is not limited to CAMs; rather, any method for extracting information of a learnt neural network or other machine learning algorithm may be used. In this example, a CAM is retrieved by attaching a global activation pooling (GAP) layer to the final convolution layer of the CNN. The GAP reduces the final activation maps having many pixels into a single (or at least fewer) representative value(s). For example, assuming the final convolution layer has k filters and the activation map of the k-th filter is Ak , the GAP is determined as the average value of all the pixels (where i,j indicates the i-th and j-th pixel) in the activation map Ak of the k-th filter according to:
-
Fig. 4 illustrates an example of a CNN with multiple convolution layers and an attached GAP layer. As seen therein, a fundus image is input to the CNN, which at a first convolutional layer applies a plurality of filters to generate a corresponding plurality of activation maps (three shown). Each of these activation maps is then applied as an input to additional convolution layers. At the final convolution layer, a plurality of filters is again applied to generate a corresponding plurality of activation maps (five shown, identified as A1 - A4 and Ak ) that are used to determine a corresponding plurality of GAPs (for example, according to Equation 1). Collectively, the GAPs are used to determine probabilities of whether the input image is of a healthy or a sick eye. - According to one example, the probability that the input image is of a sick subject is calculated according to:
where and are weights connecting the classification and the different GAPs Gk. For example, if Gk indicates the presence of drusen, could be negative and could be positive. The weights may be randomly initialized and adjusted during training of the machine learning system such that zhealthy is higher for healthy training images and zsick is higher for diseased training images. - Finally, the CAM can be calculated according to:
where C i,j indicates the likelihood that a pixel (i, j) is part of a lesion. In some embodiments, C i,j can further be rescaled to be ∈ [0,1]. By setting a threshold corresponding to a degree of likelihood that a pixel contains a lesion, individual ROIs can be identified. In other words, an ROI of the eye corresponding to a particular pixel of an input image could be identified where C i,j for that pixel is greater than the threshold. - For example,
Fig. 5 is a heat map indicating detected abnormality regions and probabilities of an abnormality of the CAM as superimposed on an input fundus image. The heat maps herein are images of the eye (e.g., the horizontal, surface, or en face images) having a color, shade, hue, or the like corresponding to a probability that the region includes an abnormality. This produces a resultant image of the eye whereby the probability of an abnormality in any particular portion of the eye is represented as a color, shade, hue, or the like in the image.Regions 500 each represent a detected abnormal region, and region 510 (having a darker shade) represents a higher probability of abnormality thanregions 500. Of course, the heat maps can be in color such that, for example, color contours highlight the detected abnormal regions and indicate a probability where a color transition from blue to red may indicate an increase in abnormality probability. In the example ofFig. 5 , a threshold of 0.4 (where Ci,j > 0.4) was used to identify theregions 500. -
Figs. 6-8 further illustrate example CAM heat maps for various retinopathies. Notably,Fig. 6 is a heat map for an image of an eye having hypertensive and arteriosclerotic retinopathies. Therein, the CAM overlay indicates 600 and 604 as likely abnormal with regions therein 602 and 606 as having the highest probability of an abnormality. Portions of the image in these high-regions 602 and 606 are enlarged, where the structural abnormality can be visually confirmed. Similarly,probability regions Fig. 7 is a heat map for an image of an eye having micro-aneurysms. The CAM map indicatedregion 700 as being abnormal, and therein,region 702 as a having a high probability of being the location of the abnormality. A portion of high-probability region 702 is enlarged 704, which shows visual confirmation of the retinopathy.Fig. 8 is a heat map for an image of an eye having a background diabetic retinopathy. Again, the CAM map indicatedregions 800 as likely containing abnormalities, withregion 802 having the highest probability of having an abnormality. Aportion 804 of high-probability region 802 is enlarged and visually confirms that the abnormality exists. - Using these CAMs and corresponding identified ROIs, a second image can be generated in and/or around the identified ROIs. The second image can be generated by a scan with a second modality (e.g., an OCT scan) that provides more detailed imaging, analysis, and information about the retinopathy. For example, an OCT scan may provide a 3D imaging volume at high resolutions so that the internal structure of retinal tissue may be analyzed, whereas the initial en-face image only images the surface of the structure. In still other examples, the second image can be an OCT-angiography (OCTA) image, visual field test results, fluorescent angiography, or fluorescent angiography fundus image. Examples of the application of CAMs according to the present disclosure are illustrated in
Figs. 9-11 . - According to the example application of
Fig. 9 , a horizontal image of the retina is taken 902 after the imaging modality used to take the image has been automatically positioned and focused 900 on the retina. Herein, a "horizontal" image means a surface or en-face image of the object being imaged, (e.g., the retina). Such an image may be taken with, for example, a fundus camera (color or infrared), scanning laser ophthalmoscope (SLO), or be a surface image derived from a 3D-OCT scan. Of course, other modalities and techniques may be used for horizontal/surface images, and the above examples are not limiting. Then, ROIs are identified from thehorizontal images 904 using the above-described non-fully-supervised machine learning and CAMs. Based on these identifications, OCT imaging and measurement is performed 906 on portions of the retina corresponding to the identified ROI locations of the horizontal image. This second imaging of the ROIs may be performed automatically (e.g., OCT imaging may be automatically controlled upon determination of the ROI) or manually instituted by a user. The data from the second image (OCT imaging) and measurements and/or the horizontal imaging is finally reported to a user and stored 908 for future analysis or review. This and other data derived from the method can also be stored, analyzed, and/or reported, for example, in any form of memory, as part of a database or the like (e.g., for future analysis or normative comparisons). The reports may include any of the images, heat maps/CAMs, identification of possible disease/retinopathy, and the like. - The application method of
Fig. 10 is similar to that ofFig. 9 , however, a 3D OCT volume of an eye is initially taken and used to obtain the horizontal image for identifying ROIs. According to this example, a second imaging scan need not be performed because all of the relevant data is captured in the initial 3D OCT volume. More particularly, the OCT imaging modality is initially positioned and focused 1000, and then the 3D OCT volume is acquired 1002. From the 3D OCT volume, a horizontal image is obtained 1004. The horizontal image may be obtained by any technique, for example, flattening the volume along a depth dimension by averaging the pixel values across a relevant depth (Z) at a particular X-Y location of the volume. Again, ROIs are identified 1004 using machine learning and CAMs. The locations of the ROIs are then translated to the original3D OCT volume 1006 so that the relevant volumetric data corresponding to the ROI can be extracted and/or otherwise highlighted 1008. All of the information including the entire 3D OCT image data can also be stored, analyzed, and/or reported; or, alternatively, the remainder of 3D image data not associated with the ROIs can be otherwise discarded. - The identified ROIs can also be useful 3D OCT scans are performed subsequently to the ROI identification. This is because the horizontal resolution of 3D OCT volumes is inversely proportional to the scan area. Thus, ROIs can guide future OCT scans at higher resolution in the most relevant regions by limiting the scan area to those most relevant regions. In other words, ROIs derived from an initial survey OCT scan that covers a large area, or similar horizontal images from large area scans from a different imaging modality, can be used to generate higher resolution scans in and around the identified ROIs by limiting the scan area. Additionally, instead of overwhelming users with a large 3D volume of data, B-scans can be selected from the ROIs that highlight an anomaly therein.
-
Fig. 11 illustrates a third application in accordance with the above. The application ofFig. 11 is similar to that ofFig. 8 , however, a 3D OCT survey image covering a large area of an eye is initially taken 1102 (after automatic positioning and focusing 1100) and used to obtain the horizontal image of theretina 1104. After identifyingROIs 1106, denser (or higher resolution) 3D OCT images are taken 1108 of the retina at locations corresponding to the ROIs to form the second images. Such denser images can reveal finer and more granular details of the tissue and can better support a particular diagnosis or disease identification, and better aid in analyzing progression of a disease. As above, the dense 3D OCT images, and/or survey image are stored, analyzed, and/or reported 1110. - In view of the above, an example system corresponding to the disclosure herein is schematically illustrated in
Fig. 12 and comprises afirst imaging modality 1200 that is capable of generating a horizontal image, asecond imaging modality 1202 capable of generating images of regions of interest identified in the horizontal image, and acomputer 1204 having aprocessor 1206, or the like configured to automatically identify the regions of interest in the horizontal image according to the above method. In view of this, computer further includes at least one processor (e.g., a central processing unit CPU, graphics processing unit GPU, or the like) that is capable of machine learning with, for example, the above-described CNN that forms amachine learning system 1212. The processor of themachine learning system 1212 may be separate from or integrated with theprocessor 1206 of thecomputer 1204. The computer could also be configured with aninput interface 1210 to receive input images from a user, or directly from the first or second imaging modalities; and anoutput interface 1210 such as a display to output the images taken, and the data collected to the user, or to directly send the ROI information to the second imaging modality. For example, these outputs may be the raw CAM or CNN data, heat maps, and the like. The system may also includememory 1208, such as RAM, ROM, flash memory, hard disks, and the like for storing the images and associated data. Of course, the first and second modalities may be the same (and comprised of common hardware features), for example, if the horizontal image and ROI image data are both (or come from) 3D OCT volume data sets collected from a single scan (as in the embodiment ofFig. 10 ). Similarly, depending on the embodiment, theprocessor 1206,memory 1208,computer 1204, and/or the like may be integrated with the imaging modalities (or lone modality) 1200, 1202, or wholly separate and simply supplied with imaging data to be analyzed. The elements of thecomputer 1204 may also be fully integrated into a single device, or separated as multiple devices, for example, if themachine learning system 1212 is embodied on a separate computer device. - The above system and methods have been tested using public datasets (e.g., publically available retinal image sets from the Structured Analysis of the Retina (STARE) Project) to characterize the performance of different types of machine learning models and configurations (e.g., how many layers are trainable and how the ROIs are extracted). The tests were performed on a computer with a Core i7 CPU and Titan Xp GPU.
- Table 1 below illustrates the specificity and sensitivity for various configurations of models and trainable convolution layers using roughly 400 fundus images having a resolution of 500×500 for training, with 39 possible retinopathy manifestations. Conv(X) in the table refers to a convolution layer with X number of filters.
Table 1. Example machine learning configurations Model Configuration # of trainable convolution layers Specificity Sensitivity Xception Add an additional Conv(20) to the end of the model 1 87% 30% Replace the final convolution layer with 1 80% 47% Conv(20) 2 93% 37% Inception (V3) Add an additional Conv(2048) layer to the end of the model 1 81% 34% ResNet50 Add an additional Conv(20) layer after the final activation layer 1 88% 47% All layers 83% 48% Add an additional Conv(20) layer before the final activation layer 1 91% 44% All layers 92% 14% InceptionResNetV2 Add an additional Conv(20) to the end of the model 1 87% 63% ∼20% of all layers 79% 41% MobileNet Add an additional Conv(1024) to the end of the model All layers 83% 31% Replace the final convolution layer with Conv(1024) 1 62% 68% All layers 81% 34% VGG16 Add an additional Conv(512) to the end of the model 1 93% 91% 4 86% 98% 7 Does not converge Replace the final convolution layer with Conv(512) 1 92% 97% 3 84% 100% 6 Does not converge VGG19* Add an additional Conv(512) to the end of the model 1 96% 92% 5 86% 100% 9 Does not converge Replace the final convolution layer with Conv(512)* 1 86% 100% 4* 88%* 98%* 8 82% 100% - As can be seen from the table, the training models achieved a good sensitivity and specificity. Thus, whereas previous machine learning studies trained and utilized one model for one type of disease, the systems and methods disclosed herein are capable of achieving high sensitivity and specificity while utilizing single models for identifying 39 different retinopathy manifestations. This success is possible with a variety of different models. While 39 retinopathy manifestations were tested, a more complex dataset (one with more retinopathy manifestations) could be used to provide the high sensitivity and specificity with more retinopathies. Thus there is no limit to the number of retinopathies to which the present disclosure can be applied. Of course the aspects of the present disclosure may be used with other models, machine learning algorithms, and methods of extracting information from those models, including those designed specifically for use with the present disclosure.
- It is noted that example images shown in
Figs. 4-8 were formed from the VGG19 model, where the final convolution layer was replaced with Conv(512), with four trainable convolution layers. This configuration is identified with an asterisk in the above table.
The present invention may also relate to the following embodiments: -
Embodiment 1. An imaging method, comprising:- generating a horizontal image of an object;
- automatically identifying a region of interest (ROI) of the object based on the horizontal image with a non-fully-supervised machine learning system; and
- generating a second image of the object within the identified ROI, wherein the second image comprises depth information of the object.
-
Embodiment 2. The method ofembodiment 1, wherein the horizontal image is a color fundus image; an infrared fundus image; a scanning laser ophthalmoscope (SLO) image; or is derived from 3D optical coherence tomography (OCT) scan data. - Embodiment 3. The method of any one of the preceding embodiments, wherein the second image is an OCT image.
-
Embodiment 4. The method of any one of the preceding embodiments, wherein the horizontal image is derived from 3D optical coherence tomography (OCT) scan data, and the second image is an OCT image generated by extracting a portion of the 3D OCT scan data corresponding to the identified ROI. - Embodiment 5. The method of
embodiment 4, further comprising discarding portions of the 3D OCT scan data that do not correspond to the identified ROI. - Embodiment 6. The method of any one of the preceding embodiments, further comprising displaying the second image.
- Embodiment 7. The method of any one of the preceding embodiments, further comprising determining probabilities of identified regions of interest, the probabilities indicating a likelihood that the region of interest represents an abnormality of the object as determined by the non-fully-supervised machine learning system.
-
Embodiment 8. The method of embodiment 7, further comprising displaying a heat map of the probabilities overlaid on the horizontal image. - Embodiment 9. The method of any one of the preceding embodiments, wherein the second image is generated from a plurality of B-scans of the region of interest.
- Embodiment 10. The method of any one of the preceding embodiments, wherein the horizontal image is derived from a 3D survey image and the second image has a greater density than the horizontal image.
- Embodiment 11. The method of any one of the preceding embodiments, wherein the horizontal image is a 3D survey image, and the second image is a 3D optical coherence tomography (OCT) image taken of the identified ROI.
- Embodiment 12. The method of any one of the preceding embodiments, further comprising only storing data corresponding to the second image, or discarding data corresponding to the horizontal image that is not associated with the identified ROI.
- Embodiment 13. The method of any one of the preceding embodiments, wherein the object is an eye.
- Embodiment 14. The method of any one of the preceding embodiments, wherein the non-fully-supervised machine learning system comprises a convolutional neural network.
- Embodiment 15. The method of any one of the preceding embodiments, wherein the ROI is identified by obtaining a class activation map of the non-fully-supervised machine learning system.
- Embodiment 16. A method of image analysis with a trained non-fully-supervised machine learning system, comprising:
- receiving a horizontal image of an object from a subject;
- identifying an abnormality of the object as an output of the trained non-fully-supervised machine learning system based on the received horizontal image;
- extracting information of the trained non-fully-supervised machine learning used to identify the abnormality;
- identifying a region of interest within the horizontal image as a region of the horizontal image that contributed to the identification of the abnormality,
- wherein the non-fully-supervised machine learning system is trained with a plurality of horizontal images of the object from different subjects to identify the abnormality of the object.
- Embodiment 17. The method of embodiment 16, wherein the trained non-fully-supervised machine learning system is a convolutional neural network.
- Embodiment 18. The method of embodiment 16 or 17, wherein the abnormality is a retinopathy disorder.
- Embodiment 19. The method of any one of the embodiments 16 to 18, wherein the information of the trained non-fully-supervised machine learning system is extracted by determining class activation maps.
- Embodiment 20. The method of embodiment 19, wherein the region of interest is identified by comparing pixel values of the determined class activation maps to a predetermined threshold.
Claims (16)
- An imaging method, comprising:generating a horizontal image of an object;identifying an abnormality of the object as an output of a trained non-fully-supervised machine learning system to which the generated horizontal image was input;extracting information of the trained non-fully-supervised machine learning used to identify the abnormality;identifying a region of interest (ROI) of the object as a region of the horizontal image that contributed to the identification of the abnormality; andgenerating a second image of the object within the identified ROI, wherein the second image comprises depth information of the object,wherein the non-fully-supervised machine learning system is trained to identify the abnormality of the object with a plurality of horizontal images of the object from different subjects.
- The method of claim 1, wherein the horizontal image is a color fundus image; an infrared fundus image; a scanning laser ophthalmoscope (SLO) image; or is derived from 3D optical coherence tomography (OCT) scan data.
- The method of any of the above claims, wherein the second image is an OCT image.
- The method of any of the above claims, wherein the horizontal image is derived from 3D optical coherence tomography (OCT) scan data, and the second image is an OCT image generated by extracting a portion of the 3D OCT scan data corresponding to the identified ROI.
- The method of any of the above claims, further comprising displaying the second image.
- The method of any of the above claims, further comprising determining probabilities of identified ROIs, each probability indicating a likelihood that a corresponding ROI represents the abnormality of the object as determined by the non-fully-supervised machine learning system.
- The method of claim 6, further comprising displaying a heat map of the probabilities overlaid on the horizontal image.
- The method of any of the above claims, wherein the horizontal image is derived from a 3D survey image and the second image has a greater density than the horizontal image.
- The method of any of the above claims, wherein the horizontal image is a 3D survey image, and the second image is a 3D optical coherence tomography (OCT) image of the identified ROI.
- The method of any of the above claims, further comprising only storing data corresponding to the second image, or discarding data corresponding to the horizontal image that is not associated with the identified ROI.
- The method of any of the above claims, wherein the object is an eye.
- The method of any of the above claims, wherein the non-fully-supervised machine learning system comprises a convolutional neural network.
- The method of any of the above claims, wherein the ROI is identified by obtaining a class activation map of the non-fully-supervised machine learning system and comparing pixel values of the determined class activation maps to a predetermined threshold.
- A system comprising:a processor configured to execute the method of any of the above claims;a first imaging system for generating the horizontal image; anda second imaging system for generating the second image.
- The system of claim 14, wherein the first and second imaging systems are integrated.
- The system of claim 14 or 15, wherein the first and second imaging systems are OCT imaging systems.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| DE18248134.1T DE18248134T1 (en) | 2017-12-28 | 2018-12-28 | MASCHINELL LEARNING, GUIDED PICTURE SYSTEM |
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US201762611352P | 2017-12-28 | 2017-12-28 | |
| US16/212,027 US11132797B2 (en) | 2017-12-28 | 2018-12-06 | Automatically identifying regions of interest of an object from horizontal images using a machine learning guided imaging system |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| EP3510917A1 true EP3510917A1 (en) | 2019-07-17 |
Family
ID=65010460
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP18248134.1A Pending EP3510917A1 (en) | 2017-12-28 | 2018-12-28 | Machine learning guided imaging system |
Country Status (4)
| Country | Link |
|---|---|
| US (2) | US11132797B2 (en) |
| EP (1) | EP3510917A1 (en) |
| JP (2) | JP2019118814A (en) |
| DE (1) | DE18248134T1 (en) |
Cited By (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP4119036A1 (en) * | 2021-07-12 | 2023-01-18 | Welch Allyn, INC. | Retinal vital sign assessment |
| JP7572316B2 (en) | 2017-12-28 | 2024-10-23 | 株式会社トプコン | Machine learning guided photography system |
Families Citing this family (40)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| KR102061408B1 (en) * | 2017-03-24 | 2019-12-31 | (주)제이엘케이인스펙션 | Apparatus and method for analyzing images using semi 3d deep neural network |
| US11200665B2 (en) * | 2017-08-02 | 2021-12-14 | Shanghai Sixth People's Hospital | Fundus image processing method, computer apparatus, and storage medium |
| CN108615051B (en) * | 2018-04-13 | 2020-09-15 | 博众精工科技股份有限公司 | Diabetic retinal image classification method and system based on deep learning |
| US11491350B2 (en) * | 2018-05-30 | 2022-11-08 | Siemens Healthcare Gmbh | Decision support system for individualizing radiotherapy dose |
| US11756667B2 (en) | 2018-05-30 | 2023-09-12 | Siemens Healthcare Gmbh | Decision support system for medical therapy planning |
| EP3847618A1 (en) * | 2018-09-12 | 2021-07-14 | Auckland Uniservices Limited | Methods and systems for ocular imaging, diagnosis and prognosis |
| US20200193156A1 (en) | 2018-12-12 | 2020-06-18 | Tesseract Health, Inc. | Biometric identification techniques |
| JP7302183B2 (en) * | 2019-01-31 | 2023-07-04 | 株式会社ニデック | Ophthalmic image processing device and ophthalmic image processing program |
| JP7302184B2 (en) * | 2019-01-31 | 2023-07-04 | 株式会社ニデック | Ophthalmic image processing device and ophthalmic image processing program |
| US11737665B2 (en) | 2019-06-21 | 2023-08-29 | Tesseract Health, Inc. | Multi-modal eye imaging with shared optical path |
| US20200397287A1 (en) * | 2019-06-21 | 2020-12-24 | Tesseract Health, Inc. | Multi-modal eye imaging applications |
| EP4006914A4 (en) * | 2019-07-31 | 2023-08-16 | Nikon Corporation | INFORMATION PROCESSING SYSTEM, INFORMATION PROCESSING DEVICE, IMAGE CAPTURE DEVICE, INFORMATION PROCESSING METHOD, IMAGE CAPTURE METHOD AND PROGRAM |
| EP4023143A4 (en) * | 2019-08-30 | 2023-08-30 | Canon Kabushiki Kaisha | INFORMATION PROCESSING DEVICE, INFORMATION PROCESSING METHOD, INFORMATION PROCESSING SYSTEM AND PROGRAM |
| JP7596092B2 (en) * | 2019-08-30 | 2024-12-09 | キヤノン株式会社 | Information processing device, information processing method, information processing system, and program |
| CN110969191B (en) * | 2019-11-07 | 2022-10-25 | 吉林大学 | Glaucoma prevalence probability prediction method based on similarity maintenance metric learning method |
| WO2021108458A1 (en) * | 2019-11-25 | 2021-06-03 | Broadspot Imaging Corp | Choroidal imaging |
| JP7332463B2 (en) * | 2019-12-25 | 2023-08-23 | キヤノン株式会社 | Control device, optical coherence tomography device, control method for optical coherence tomography device, and program |
| AU2021209954A1 (en) * | 2020-01-23 | 2022-09-01 | Retispec Inc. | Systems and methods for disease diagnosis |
| US11508061B2 (en) * | 2020-02-20 | 2022-11-22 | Siemens Healthcare Gmbh | Medical image segmentation with uncertainty estimation |
| CN111369528B (en) * | 2020-03-03 | 2022-09-09 | 重庆理工大学 | A deep convolutional network-based method for marking stenotic regions in coronary angiography images |
| US20210304001A1 (en) * | 2020-03-30 | 2021-09-30 | Google Llc | Multi-head neural network model to simultaneously predict multiple physiological signals from facial RGB video |
| JP7413147B2 (en) * | 2020-05-21 | 2024-01-15 | キヤノン株式会社 | Image processing device, image processing method, and program |
| LT6911B (en) | 2020-07-17 | 2022-05-10 | Vilniaus Universitetas | Analysis system for images and data from ultrasound examination and ultrasound examination with contrast for early diagnosis of pancreatic pathologies by automatic means |
| JP6887199B1 (en) * | 2020-08-04 | 2021-06-16 | 株式会社オプティム | Computer system, dataset creation method and program |
| JP7687669B2 (en) * | 2020-08-19 | 2025-06-03 | 国立研究開発法人農業・食品産業技術総合研究機構 | Leaf temperature acquisition device, crop cultivation system, leaf temperature acquisition method, program for acquiring leaf temperature, and crop cultivation method |
| US11532147B2 (en) * | 2020-09-25 | 2022-12-20 | Microsoft Technology Licensing, Llc | Diagnostic tool for deep learning similarity models |
| WO2022072513A1 (en) | 2020-09-30 | 2022-04-07 | Ai-Ris LLC | Retinal imaging system |
| JP7785294B2 (en) * | 2020-11-04 | 2025-12-15 | 株式会社トプコン | Ophthalmological information processing device, ophthalmological device, ophthalmological information processing method, and program |
| JP7627564B2 (en) * | 2020-11-05 | 2025-02-06 | グローリー株式会社 | IMAGE PROCESSING DEVICE, IMAGE PROCESSING METHOD, LEARNING MODEL GENERATION DEVICE, LEARNING MODEL PRODUCTION METHOD, LEARNED MODEL, AND PROGRAM |
| JP7644329B2 (en) * | 2020-12-08 | 2025-03-12 | キヤノンマーケティングジャパン株式会社 | Information processing device, information processing method, and program |
| KR20220085481A (en) | 2020-12-15 | 2022-06-22 | 삼성전자주식회사 | Method and apparatus of image processing |
| JP7706057B2 (en) * | 2021-03-10 | 2025-07-11 | 株式会社ニデック | Ophthalmic image processing device, ophthalmic image processing program, and ophthalmic image capturing device |
| JP7714916B2 (en) * | 2021-06-03 | 2025-07-30 | 株式会社ニデック | OCT device and imaging control program |
| JP2023128693A (en) * | 2022-03-04 | 2023-09-14 | コニカミノルタ株式会社 | Information processing device, analysis program, and analysis method |
| CN114677503A (en) * | 2022-03-10 | 2022-06-28 | 上海联影医疗科技股份有限公司 | Region of interest detection method, apparatus, computer equipment and storage medium |
| CA3257993A1 (en) * | 2022-07-13 | 2024-01-18 | Alcon Inc. | Machine learning to assess the clinical significance of vitreous floaters |
| KR20240062454A (en) * | 2022-11-01 | 2024-05-09 | 삼성전자주식회사 | Method and apparatus for extracting the region of interest |
| CN118736027A (en) * | 2023-03-29 | 2024-10-01 | 上海联影医疗科技股份有限公司 | Medical image optimization method, device, equipment, storage medium and program product |
| EP4523607A1 (en) * | 2023-09-14 | 2025-03-19 | Optos PLC | A multi-modality ophthalmic imaging system and a method of imaging a patient s eye |
| US20260094709A1 (en) * | 2024-09-27 | 2026-04-02 | Stephen McNutt | Integrated mixed reality visualization for diagnostic imaging and data mapping |
Citations (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20140276025A1 (en) * | 2013-03-14 | 2014-09-18 | Carl Zeiss Meditec, Inc. | Multimodal integration of ocular data acquisition and analysis |
| US20140314288A1 (en) * | 2013-04-17 | 2014-10-23 | Keshab K. Parhi | Method and apparatus to detect lesions of diabetic retinopathy in fundus images |
| CN107423571A (en) * | 2017-05-04 | 2017-12-01 | 深圳硅基仿生科技有限公司 | Diabetic retinopathy recognition system based on fundus images |
Family Cites Families (17)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2008062528A1 (en) * | 2006-11-24 | 2008-05-29 | Nidek Co., Ltd. | Fundus image analyzer |
| US8737703B2 (en) | 2008-01-16 | 2014-05-27 | The Charles Stark Draper Laboratory, Inc. | Systems and methods for detecting retinal abnormalities |
| US20090309874A1 (en) * | 2008-06-11 | 2009-12-17 | Siemens Medical Solutions Usa, Inc. | Method for Display of Pre-Rendered Computer Aided Diagnosis Results |
| JP4819851B2 (en) * | 2008-07-31 | 2011-11-24 | キヤノン株式会社 | Diagnosis support apparatus and method, program, and recording medium |
| WO2012174495A2 (en) * | 2011-06-17 | 2012-12-20 | Carnegie Mellon University | Physics based image processing and evaluation process of perfusion images from radiology imaging |
| JP6071331B2 (en) * | 2012-08-27 | 2017-02-01 | キヤノン株式会社 | Image processing apparatus and image processing method |
| US9107610B2 (en) | 2012-11-30 | 2015-08-18 | Kabushiki Kaisha Topcon | Optic neuropathy detection with three-dimensional optical coherence tomography |
| JP6241040B2 (en) * | 2013-01-23 | 2017-12-06 | 株式会社ニデック | Ophthalmic analysis apparatus and ophthalmic analysis program |
| US9179834B2 (en) * | 2013-02-01 | 2015-11-10 | Kabushiki Kaisha Topcon | Attenuation-based optic neuropathy detection with three-dimensional optical coherence tomography |
| US9241626B2 (en) | 2013-03-14 | 2016-01-26 | Carl Zeiss Meditec, Inc. | Systems and methods for improved acquisition of ophthalmic optical coherence tomography data |
| US10115194B2 (en) * | 2015-04-06 | 2018-10-30 | IDx, LLC | Systems and methods for feature detection in retinal images |
| US10304198B2 (en) * | 2016-09-26 | 2019-05-28 | Siemens Healthcare Gmbh | Automatic medical image retrieval |
| KR101879207B1 (en) * | 2016-11-22 | 2018-07-17 | 주식회사 루닛 | Method and Apparatus for Recognizing Objects in a Weakly Supervised Learning Manner |
| US10660576B2 (en) * | 2017-01-30 | 2020-05-26 | Cognizant Technology Solutions India Pvt. Ltd. | System and method for detecting retinopathy |
| US10152571B1 (en) * | 2017-05-25 | 2018-12-11 | Enlitic, Inc. | Chest x-ray differential diagnosis system |
| CN108230294B (en) * | 2017-06-14 | 2020-09-29 | 北京市商汤科技开发有限公司 | Image detection method, image detection device, electronic equipment and storage medium |
| US11132797B2 (en) | 2017-12-28 | 2021-09-28 | Topcon Corporation | Automatically identifying regions of interest of an object from horizontal images using a machine learning guided imaging system |
-
2018
- 2018-12-06 US US16/212,027 patent/US11132797B2/en active Active
- 2018-12-25 JP JP2018241158A patent/JP2019118814A/en active Pending
- 2018-12-28 DE DE18248134.1T patent/DE18248134T1/en active Pending
- 2018-12-28 EP EP18248134.1A patent/EP3510917A1/en active Pending
-
2021
- 2021-06-29 JP JP2021107210A patent/JP7572316B2/en active Active
- 2021-09-13 US US17/447,465 patent/US20210407088A1/en active Pending
Patent Citations (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20140276025A1 (en) * | 2013-03-14 | 2014-09-18 | Carl Zeiss Meditec, Inc. | Multimodal integration of ocular data acquisition and analysis |
| US20140314288A1 (en) * | 2013-04-17 | 2014-10-23 | Keshab K. Parhi | Method and apparatus to detect lesions of diabetic retinopathy in fundus images |
| CN107423571A (en) * | 2017-05-04 | 2017-12-01 | 深圳硅基仿生科技有限公司 | Diabetic retinopathy recognition system based on fundus images |
Non-Patent Citations (1)
| Title |
|---|
| CARSON LAM ET AL: "Imaging Retinal Lesion Detection With Deep Learning Using Image Patches", 1 January 2018 (2018-01-01), XP055594611, Retrieved from the Internet <URL:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC5788045/pdf/i1552-5783-59-1-590.pdf> [retrieved on 20190606] * |
Cited By (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP7572316B2 (en) | 2017-12-28 | 2024-10-23 | 株式会社トプコン | Machine learning guided photography system |
| EP4119036A1 (en) * | 2021-07-12 | 2023-01-18 | Welch Allyn, INC. | Retinal vital sign assessment |
| US12533022B2 (en) | 2021-07-12 | 2026-01-27 | Welch Allyn, Inc. | Retinal vital sign assessment |
Also Published As
| Publication number | Publication date |
|---|---|
| US20210407088A1 (en) | 2021-12-30 |
| US20190206054A1 (en) | 2019-07-04 |
| JP2019118814A (en) | 2019-07-22 |
| JP7572316B2 (en) | 2024-10-23 |
| JP2021154159A (en) | 2021-10-07 |
| DE18248134T1 (en) | 2019-10-02 |
| US11132797B2 (en) | 2021-09-28 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US11132797B2 (en) | Automatically identifying regions of interest of an object from horizontal images using a machine learning guided imaging system | |
| EP3671536B1 (en) | Detection of pathologies in ocular images | |
| AU2018289501B2 (en) | Segmentation of retinal blood vessels in optical coherence tomography angiography images | |
| EP3543880B1 (en) | Stroke diagnosis and prognosis system | |
| Niemeijer et al. | Retinopathy online challenge: automatic detection of microaneurysms in digital color fundus photographs | |
| US20210104313A1 (en) | Medical image processing apparatus, medical image processing method and computer-readable medium | |
| US8556424B2 (en) | Image processing apparatus, image processing method, and program | |
| US20230394689A1 (en) | Method and system for imaging a biological tissue | |
| EP3719807B1 (en) | Predicting a pathological condition from a medical image | |
| US20210224957A1 (en) | Medical image processing apparatus, medical image processing method and computer-readable medium | |
| US9898818B2 (en) | Automated measurement of changes in retinal, retinal pigment epithelial, or choroidal disease | |
| CN110717487A (en) | Methods and systems for identifying cerebrovascular abnormalities | |
| JP5631339B2 (en) | Image processing apparatus, image processing method, ophthalmic apparatus, ophthalmic system, and computer program | |
| US20240296555A1 (en) | Machine learning classification of retinal atrophy in oct scans | |
| US12573032B2 (en) | System and methods of predicting Parkinson's disease based on retinal images using machine learning | |
| JP7475968B2 (en) | Medical image processing device, method and program | |
| US7961923B2 (en) | Method for detection and visional enhancement of blood vessels and pulmonary emboli | |
| RU120799U1 (en) | INTEREST AREA SEARCH SYSTEM IN THREE-DIMENSIONAL MEDICAL IMAGES | |
| Ramya et al. | Diabetic retinopathy analysis using machine learning | |
| Azmi et al. | Brain anatomical variations among Malaysian population |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| PUAI | Public reference made under article 153(3) epc to a published international application that has entered the european phase |
Free format text: ORIGINAL CODE: 0009012 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE APPLICATION HAS BEEN PUBLISHED |
|
| AK | Designated contracting states |
Kind code of ref document: A1 Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC MK MT NL NO PL PT RO RS SE SI SK SM TR |
|
| AX | Request for extension of the european patent |
Extension state: BA ME |
|
| REG | Reference to a national code |
Ref country code: DE Ref legal event code: R210 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: REQUEST FOR EXAMINATION WAS MADE |
|
| 17P | Request for examination filed |
Effective date: 20200117 |
|
| RBV | Designated contracting states (corrected) |
Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC MK MT NL NO PL PT RO RS SE SI SK SM TR |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: EXAMINATION IS IN PROGRESS |
|
| 17Q | First examination report despatched |
Effective date: 20211022 |
|
| REG | Reference to a national code |
Ref country code: DE Ref legal event code: R079 Free format text: PREVIOUS MAIN CLASS: A61B0003100000 Ipc: G06F0018240000 |
|
| GRAP | Despatch of communication of intention to grant a patent |
Free format text: ORIGINAL CODE: EPIDOSNIGR1 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: GRANT OF PATENT IS INTENDED |
|
| RIC1 | Information provided on ipc code assigned before grant |
Ipc: G06F 18/24 20230101AFI20260310BHEP Ipc: G06V 10/44 20220101ALI20260310BHEP Ipc: G06V 10/764 20220101ALI20260310BHEP Ipc: A61B 3/10 20060101ALI20260310BHEP Ipc: A61B 3/12 20060101ALI20260310BHEP Ipc: G06T 7/00 20170101ALI20260310BHEP Ipc: G06F 18/2413 20230101ALI20260310BHEP Ipc: G06V 10/82 20220101ALI20260310BHEP |
|
| INTG | Intention to grant announced |
Effective date: 20260323 |




