US2021110196A1PendingUtilityA1

Deep Learning Network for Salient Region Identification in Images

Assignee: IBMPriority: Dec 10, 2018Filed: Dec 22, 2020Published: Apr 15, 2021
Est. expiryDec 10, 2038(~12.4 yrs left)· nominal 20-yr term from priority
G06T 12/00G06N 3/088G06V 10/26G06V 10/82G06V 10/454G06V 10/764G06F 18/24143G06N 3/042G06N 3/045G06N 3/0464G06N 3/09G06V 2201/03G06T 7/0012G06N 5/041G06T 2207/30016G06T 2207/20081G06T 2207/10081G06T 2207/20084G06N 3/08G06T 2207/10088G06T 2207/10104G06T 2207/10132G06T 11/003G06K 9/4609G06K 9/4676
64
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

Mechanisms are provided to implement a hybrid deep learning network. The hybrid deep learning network receives, from a imaging system, first input data specifying a non-annotated image. The hybrid deep learning network pre-processes the non-annotated image to generate second input data specifying a hint image and corresponding annotation data specifying salient regions of the hint image. The hybrid deep learning network processes the first input data and second input data to perform training of the hybrid deep learning network by targeting feature detection in the non-annotated image in the salient regions identified in the hint image. The trained hybrid deep learning network is used to process third input data specifying a new non-annotated image to thereby identify an object or structure in the new non-annotated image.

Claims

exact text as granted — not AI-modified
1 . A method, in a data processing system comprising at least one processor and at least one memory, wherein the at least one memory comprises instructions that are executed by the at least one processor to cause the at least one processor to implement a hybrid deep learning network, and wherein the method comprises:
 receiving, by the hybrid deep learning network, from an imaging system, first input data specifying a non-annotated image;   pre-processing, by the hybrid deep learning network, the non-annotated image to generate second input data specifying a hint image and corresponding annotation data specifying salient regions of the hint image;   processing, by the hybrid deep learning network, the first input data and second input data to perform training of the hybrid deep learning network by targeting feature detection in the non-annotated image in the salient regions identified in the hint image; and   processing, using the trained hybrid deep learning network, third input data specifying a new non-annotated image to thereby identify an object or structure in the new non-annotated image.   
     
     
         2 . The method of  claim 1 , wherein processing the first input data and second input data to perform training of the hybrid deep learning network by targeting feature detection in the non-annotated image in the salient regions identified in the hint image comprises filtering out regions of the non-annotated image that are not specified in the hint image as being salient regions. 
     
     
         3 . The method of  claim 1 , wherein the pre-processing of the non-annotated image comprises generating the hint image at least by performing, on the non-annotated image, a multi-level thresholding operation with region grouping based on one or more saliency operators. 
     
     
         4 . The method of  claim 3 , wherein the plurality of saliency operations comprise at least one image characteristic, and wherein the at least one image characteristic comprises at least one of a region size, a region location, color value, or an intensity value. 
     
     
         5 . The method of  claim 1 , wherein the pre-processing of the non-annotated image comprises applying a region size filter on regions of the non-annotated image having different tissue densities to thereby identify salient regions within the non-annotated image. 
     
     
         6 . The method of  claim 5 , wherein the pre-processing of the non-annotated image further comprises performing a color connected component grouping that identifies salient regions within the non-annotated image, where anatomical structures in the non-annotated image that have similar characteristics have similar coloring in the non-annotated image. 
     
     
         7 . The method of  claim 6 , further comprising:
 performing task specific filtering of salient regions in the pre-processed non-annotated image to identify salient regions of interest to the particular task being performed; and   generating the hint image based on the filtered salient regions, wherein the task specific filtering comprises applying task specific saliency measures indicating either positive or negative saliency for the particular task.   
     
     
         8 . The method of  claim 1 , wherein the annotation data specifies one or more contours in the non-annotated image, and one or more corresponding labels, identifying anatomical structures present in the non-annotated image. 
     
     
         9 . (canceled) 
     
     
         10 . The method of  claim 1 , wherein the imaging system is at least one of a x-ray imaging system, a sonogram imaging system, a computed tomography (CT) scan imaging system, a positron emission tomography (PET) scan imaging system, magnetic resonance image (MRI) imaging system, or an echocardiography imaging system. 
     
     
         11 . A computer program product comprising a computer readable storage medium having a computer readable program stored therein, wherein the computer readable program, when executed on a data processing system, causes the data processing system to implement a hybrid deep learning network that operates to:
 receive, from an imaging system, first input data specifying a non-annotated image;   pre-process the non-annotated image to generate second input data specifying a hint image and corresponding annotation data specifying salient regions of the hint image; and   process the first input data and second input data to perform training of the hybrid deep learning network by targeting feature detection in the non-annotated image in the salient regions identified in the hint image, and wherein the data processing system further processes, using the trained hybrid deep learning network, third input data specifying a new non-annotated image to thereby identify an object or structure in the new non-annotated image.   
     
     
         12 . The computer program product of  claim 11 , wherein the computer readable program further causes the hybrid deep learning network to process the first input data and second input data to perform training of the hybrid deep learning network by targeting feature detection in the non-annotated image in the salient regions identified in the hint image at least by filtering out regions of the non-annotated image that are not specified in the hint image as being salient regions. 
     
     
         13 . The computer program product of  claim 11 , wherein the computer readable program further causes the hybrid deep learning network to pre-process the non-annotated image at least by generating the hint image at least by performing, on the non-annotated image, a multi-level thresholding operation with region grouping based on one or more saliency operators. 
     
     
         14 . The computer program product of  claim 13 , wherein the plurality of saliency operations comprise at least one image characteristic, and wherein the at least one image characteristic comprises at least one of a region size, a region location, color value, or an intensity value. 
     
     
         15 . The computer program product of  claim 11 , wherein the computer readable program further causes the hybrid deep learning network to pre-process the non-annotated image at least by applying a region size filter on regions of the non-annotated image having different tissue densities to thereby identify salient regions within the non-annotated image. 
     
     
         16 . The computer program product of  claim 15 , wherein the computer readable program further causes the hybrid deep learning network to pre-process the non-annotated image further at least by performing a color connected component grouping that identifies salient regions within the non-annotated image, where anatomical structures in the non-annotated image that have similar characteristics have similar coloring in the non-annotated image. 
     
     
         17 . The computer program product of  claim 16 , wherein the computer readable program further causes the hybrid deep learning network to:
 perform task specific filtering of salient regions in the pre-processed non-annotated image to identify salient regions of interest to the particular task being performed; and   generate the hint image based on the filtered salient regions, wherein the task specific filtering comprises applying task specific saliency measures indicating either positive or negative saliency for the particular task.   
     
     
         18 . The computer program product of  claim 11 , wherein the annotation data specifies one or more contours in the non-annotated image, and one or more corresponding labels, identifying anatomical structures present in the non-annotated image. 
     
     
         19 . (canceled) 
     
     
         20 . A data processing system comprising:
 at least one processor; and   at least one memory coupled to the at least one processor, wherein the at least one memory comprises instructions which, when executed by the at least one processor, cause the at least one processor to implement a hybrid deep learning network that operates to:   receive, from an imaging system, first input data specifying a non-annotated image;   pre-process the non-annotated image to generate second input data specifying a hint image and corresponding annotation data specifying salient regions of the hint image; and   process the first input data and second input data to perform training of the hybrid deep learning network by targeting feature detection in the non-annotated image in the salient regions identified in the hint image, and wherein the data processing system further processes, using the trained hybrid deep learning network, third input data specifying a new non-annotated image to thereby identify an object or structure in the new non-annotated image.

Join the waitlist — get patent alerts

Track US2021110196A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.