IP Library Granted Patent US 12,293,521
Granted Patent B2
US 12,293,521 · App. 17/901,429 · Granted May 6, 2025

Apparatus and methods for image segmentation using machine learning processes

Inventors: Chung-Chi Tsai (San Diego, CA); Shubhankar Mangesh Borse (San Diego, CA); Meng-Lin Wu (San Diego, CA); Venkata Ravi Kiran Dayana (San Diego, CA); Fatih Murat Porikli (San Diego, CA); An Chen (San Diego, CA)
Assignee: QUALCOMM Incorporated
G06T7/11G06T7/74G06V30/19153G06V30/19173G06T2207/20112
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,293,521
App. No.
17/901,429
Granted
May 6, 2025
Kind
B2
Abstract

Methods, systems, and apparatuses for image segmentation are provided. For example, a computing device may obtain an image, and may apply a process to the image to generate input image feature data and input image segmentation data. Further, the computing device may obtain reference image feature data and reference image classification data for a plurality of reference images. The computing device may generate reference image segmentation data based on the reference image feature data, the reference image classification data, and the input image feature data. The computing device may further blend the input image segmentation data and the reference image segmentation data to generate blended image segmentation data. The computing device may store the blended image segmentation data within a data repository. In some examples, the computing device provides the blended image segmentation data for display.

Claims (46)

1. An apparatus comprising:

a non-transitory, machine-readable storage medium storing instructions; and

at least one processor coupled to the non-transitory, machine-readable storage medium, the at least one processor being configured to:

apply a process to an input image to generate input image feature data characterizing features of the input image, and input image segmentation data characterizing an initial segmentation of the input image;

obtain, for a plurality of reference images, reference image feature data characterizing features of the plurality of reference images, and reference image classification data characterizing classifications of the plurality of reference images;

generate reference image segmentation data based on the reference image feature data, the reference image classification data, and the input image feature data; and

combine values of the input image segmentation data with values of the reference image segmentation data to generate blended image segmentation data, the blended image segmentation data characterizing a final segmentation of the input image.

2. The apparatus of claim 1 , wherein the at least one processor is further configured to execute the instructions to:

generate, for each of the plurality of reference images, first reference image data based on the reference image feature data and the reference image classification data corresponding to each of the plurality of reference images; and

generate the reference image segmentation data based on the first reference image data for each of the plurality of reference images and the input image feature data.

3. The apparatus of claim 2 , wherein the at least one processor is further configured to execute the instructions to:

generate, for each of the plurality of reference images, second reference image data based on the first reference image data and the input image feature data; and

generate the reference image segmentation data based on the second reference image data for each of the plurality of reference images.

4. The apparatus of claim 3 , wherein the at least one processor is further configured to execute the instructions to obtain the plurality of reference images from a data repository.

5. The apparatus of claim 1 , wherein the at least one processor is further configured to execute the instructions to obtain the reference image feature data from a data repository.

6. The apparatus of claim 1 , wherein the at least one processor is further configured to execute the instructions to apply the process to the plurality of reference images to generate the reference image feature data.

7. The apparatus of claim 1 , wherein the at least one processor is further configured to execute the instructions to display the blended image segmentation data.

8. The apparatus of claim 1 , wherein the at least one processor is further configured to execute the instructions to store the blended image segmentation data within a data repository.

9. The apparatus of claim 1 , wherein the process is a trained machine learning process, and wherein the at least one processor is further configured to execute the instructions to establish a convolutional neural network to apply the trained machine learning process to the input image.

10. The apparatus of claim 9 , wherein the input image feature data is an output of an encoder of the convolutional neural network, and the input image segmentation data is an output of a decoder of the convolutional neural network.

11. The apparatus of claim 1 , wherein the at least one processor is further configured to execute the instructions to:

receive a first input from a user; and

capture the input image in response to the first input from the user.

12. The apparatus of claim 1 , wherein the apparatus is a head mounted device.

13. The apparatus of claim 1 , wherein the at least one processor is further configured to execute the instructions to apply at least one image processing operation to the input image based on the blended image segmentation data.

14. A method for image segmentation comprising:

applying a process to an input image to generate input image feature data characterizing features of the input image, and input image segmentation data characterizing an initial segmentation of the input image;

obtaining, for a plurality of reference images, reference image feature data characterizing features of the plurality of reference images, and reference image classification data characterizing classifications of the plurality of reference images;

generating reference image segmentation data based on the reference image feature data, the reference image classification data, and the input image feature data; and

combining values of the input image segmentation data with values of the reference image segmentation data to generate blended image segmentation data, the blended image segmentation data characterizing a final segmentation of the input image.

15. The method of claim 14 , comprising:

generating, for each of the plurality of reference images, first reference image data based on the reference image feature data and the reference image classification data corresponding to each of the plurality of reference images; and

generating the reference image segmentation data based on the first reference image data for each of the plurality of reference images and the input image feature data.

16. The method of claim 15 , comprising:

generating, for each of the plurality of reference images, second reference image data based on the first reference image data and the input image feature data; and

generating the reference image segmentation data based on the second reference image data for each of the plurality of reference images.

17. The method of claim 14 , comprising applying the process to the plurality of reference images to generate the reference image feature data.

18. The method of claim 14 , wherein the process is a trained machine learning process, the method comprising establishing a convolutional neural network to apply the trained machine learning process to the input image.

19. A non-transitory, machine-readable storage medium storing instructions that, when executed by at least one processor, cause the at least one processor to perform operations that include:

applying a process to an input image to generate input image feature data characterizing features of the input image, and input image segmentation data characterizing an initial segmentation of the input image;

obtaining, for a plurality of reference images, reference image feature data characterizing features of the plurality of reference images, and reference image classification data characterizing classifications of the plurality of reference images;

generating reference image segmentation data based on the reference image feature data, the reference image classification data, and the input image feature data; and

combining values of the input image segmentation data with values of the reference image segmentation data to generate blended image segmentation data, the blended image segmentation data characterizing a final segmentation of the input image.

20. The non-transitory, machine-readable storage medium of claim 19 , wherein the instructions, when executed by the at least one processor, cause the at least one processor to perform additional operations that include:

generating, for each of the plurality of reference images, first reference image data based on the reference image feature data and the reference image classification data corresponding to each of the plurality of reference images; and

generating the reference image segmentation data based on the first reference image data for each of the plurality of reference images and the input image feature data.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 13, 2022
From: TSAI, CHUNG-CHI; BORSE, SHUBHANKAR MANGESH; WU, MENG-LIN; DAYANA, VENKATA RAVI KIRAN; PORIKLI, FATIH MURAT; CHEN, AN
To: QUALCOMM INCORPORATED
Reel/Frame 061417/0470 →
Continuity (1)
Related Publication 20240078679A1 · Mar 7, 2024
References Cited (11)
US 20100316294A1 · Perner · 2010 [cited by examiner]
US 20220292684A1 · Wang · 2022 [cited by examiner]
US 20230334318A1 · Liu · 2023 [cited by examiner]
Broomhead D.S., et al., “Radial Basis Functions, Multi-Variable Functional Interpolation and Adaptive Networks”, Royal Signals and Radar Establishment, Malvern (United Kingdom), Mar. 28, 1988, 39 Pages. [cited by applicant]
Bubeck S., et al., “A Law of Robustness for Two-Layers Neural Networks”, Conference on Learning Theory, PMLR, 2021, pp. 1-17. [cited by applicant]
Castro F.M., et al., “End-to-End Incremental Learning”, Proceedings of the European Conference on Computer Vision (ECCV), Sep. 2018, 21 Pages. [cited by applicant]
Hayes T.L., et al., “Remind Your Neural Network to Prevent Catastrophic Forgetting”, European Conference on Computer Vision, Springer, Cham, 2020, 18 Pages. [cited by applicant]
Li T., et al., “Federated Learning: Challenges, Methods, and Future Directions”, IEEE Signal Processing Magazine 37.3, May 2020, pp. 50-60. [cited by applicant]
Tsai C-C., et al., “Deep Co-Saliency Detection via Stacked Autoencoder-Enabled Fusion and Self-Trained CNNs”, IEEE Transactions on Multimedia 22.4 (2019), pp. 1-16, Apr. 2020. [cited by applicant]
Tsai C-C., et al., “Image Co-Saliency Detection and Co-Segmentation via Progressive Joint Optimization”, IEEE Transactions on Image Processing 28.1 (2018), DOI:10.1109/TIP.2018.2861217, Jul. 30, 2018, pp. 1-16. [cited by applicant]
Xiaofeng L., et al., “Data Augmentation via Latent Space Interpolation for Image Classification”, 2018 24th International Conference on Pattern Recognition (ICPR). IEEE, Aug. 20-24, 2018, pp. 728-733. [cited by applicant]