IP Library › Granted Patent US 12,443,751
Granted Patent B2
US 12,443,751 · App. 18/041,418 · Granted Oct 14, 2025

Anonymizing textual content in image training data

Inventors: Daniel Albertini (Vienna, AT); Martin Cerman (Vienna, AT); Michael Schwarz (Vienna, AT); Aniello Raffaele Patrone (Vienna, AT)
Assignee: ANYLINE GMBH
G06F21/6254G06F40/166G06T5/70G06V40/161
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,443,751
App. No.
18/041,418
Granted
Oct 14, 2025
Kind
B2
Abstract

A computer-implemented method for modifying image data, the method including: loading unmodified image data; detecting at least two alphanumeric characters in an image represented by the loaded image data; selecting one or more of the detected alphanumeric characters, wherein the number of selected alphanumeric characters is smaller than the total number of detected alphanumeric characters; modifying the loaded image data by removing one or more character sections of the loaded image data, wherein each character section corresponds to an area of a selected alphanumeric character; storing the modified image data. Also provided is a corresponding data processing system, computer program product and computer-readable storage medium.

Claims (47)

1. A computer-implemented method for modifying image data, the method comprising:

loading unmodified image data;

detecting at least two alphanumeric characters in an image represented by the loaded image data;

selecting one or more of the detected alphanumeric characters, wherein the number of selected alphanumeric characters is smaller than the total number of detected alphanumeric characters;

modifying the loaded image data by removing one or more character sections of the loaded image data, wherein each character section corresponds to an area of a selected alphanumeric character;

modifying the loaded image data by shuffling one or more character sections of the loaded image data, wherein each character section corresponds to an area of an unselected alphanumeric character; and

storing the modified image data.

2. The method of claim 1 , wherein the number of selected alphanumeric characters is approximately half the number of all detected alphanumeric characters.

3. The method of claim 1 , wherein:

detecting at least one word comprising two or more alphanumeric characters in the image represented by the loaded image data,

selecting one or more of the detected alphanumeric characters of each detected word.

4. The method of claim 3 , further comprising selecting approximately half the number of detected alphanumeric characters of each word.

5. The method of claim 1 , further comprising:

modifying the loaded image data by replacing at least one of the removed character sections of the loaded image data with a character section of the loaded image data corresponding to an area of an unselected alphanumeric character.

6. The method of claim 3 , further comprising:

modifying the loaded image data by replacing at least one of the removed character sections of the loaded image data belonging to the at least one word with a character section of the loaded image data corresponding to an area of an unselected alphanumeric character belonging to the same word as the removed character section.

7. The method of claim 3 , wherein shuffling of character sections corresponding to alphanumeric characters belonging to the at least one word is limited to shuffling within the same word.

8. The method of claim 1 , further comprising:

detecting at least one face in an image represented by the loaded image data;

modifying the loaded image data by removing one or more portrait sections of the loaded image data, wherein each portrait section corresponds to an area of a detected face.

9. The method of claim 8 , wherein removing one or more portrait sections of the loaded image data comprises replacing at least one removed portrait section with a blurred version of the same portrait section.

10. The method of claim 8 , further comprising:

detecting at least one additional face in an image represented by the loaded image data using the at least one detected face as a template;

modifying the loaded image data by removing one or more additional portrait sections of the loaded image data, wherein each additional portrait section corresponds to an area of a detected additional face.

11. The method of claim 10 , wherein removing one or more additional portrait sections of the loaded image data comprises replacing at least one additional portrait section with a blurred version of the same additional portrait section.

12. The method of claim 1 , further comprising:

detecting at least one written signature in an image represented by the loaded image data;

modifying the loaded image data by removing one or more signature sections of the loaded image data, wherein each signature section corresponds to an area of a detected written signature.

13. The method of claim 12 , wherein removing one or more signature sections of the loaded image data comprises replacing at least one removed signature section with a blurred version of the same signature section.

14. The method of claim 1 , further comprising:

detecting at least one machine-readable code in an image represented by the loaded image data;

modifying the loaded image data by removing one or more code sections of the loaded image data, wherein each code section corresponds to an area of a detected machine-readable code.

15. The method of claim 14 , wherein removing one or more code sections of the loaded image data comprises replacing at least one removed code section with a blurred version of the same code section.

16. A data processing system comprising means for carrying out the method of claim 1 .

17. A computer program product comprising instructions which, when the program is executed by a computer, cause the computer to carry out the method of claim 1 .

18. A computer-implemented method for modifying image data, the method comprising:

loading unmodified image data;

detecting at least two alphanumeric characters in an image represented by the loaded image data;

selecting one or more of the detected alphanumeric characters, wherein the number of selected alphanumeric characters is smaller than the total number of detected alphanumeric characters;

detecting at least one word comprising two or more alphanumeric characters in the image represented by the loaded image data;

selecting one or more of the detected alphanumeric characters of each detected word;

selecting approximately half the number of detected alphanumeric characters of each word;

modifying the loaded image data by removing one or more character sections of the loaded image data, wherein each character section corresponds to an area of a selected alphanumeric character; and

storing the modified image data.

19. The method of claim 18 , further comprising:

modifying the loaded image data by replacing at least one of the removed character sections of the loaded image data belonging to the at least one word with a character section of the loaded image data corresponding to an area of an unselected alphanumeric character belonging to the same word as the removed character section.

20. The method of claim 18 , wherein the number of selected alphanumeric characters is approximately half the number of all detected alphanumeric characters.

Assignments (2)
SECURITY INTEREST Recorded Jul 13, 2023
From: ANYLINE GMBH
To: ATEMPO GROWTH I
Reel/Frame 064565/0233 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 3, 2023
From: ALBERTINI, DANIEL; CERMAN, MARTIN; SCHWARZ, MICHAEL; PATRONE, ANIELLO RAFFAELE
To: ANYLINE GMBH
Reel/Frame 062865/0312 →
Priority Claims (1)
EP 20190903 · Aug 13, 2020 · regional
Continuity (1)
Related Publication 20240028763A1 · Jan 25, 2024
References Cited (16)
US 20050248808A1 · Ma et al. · 2005 [cited by applicant]
US 20140136941A1 · Avrahami · 2014 [cited by examiner]
US 20140355069A1 · Caton et al. · 2014 [cited by applicant]
US 20200065521A1 · Durvasula et al. · 2020 [cited by applicant]
US 20200244626A1 · Kwon · 2020 [cited by examiner]
US 20210248270A1 · Durvasula et al. · 2021 [cited by applicant]
EP 3188058A1 · 2017 [cited by applicant]
EP 3451209A1 · 2019 [cited by applicant]
EP 3614291A1 · 2020 [cited by applicant]
JP 2002149638A · 2002 [cited by applicant]
JP 2015085533A · 2015 [cited by applicant]
JP 2015222460A · 2015 [cited by applicant]
Baek, Y. et al., “Character Region Awareness for Text Detection”, arXiv:1904.01941 [cs.CV], Apr. 3, 2019. [cited by applicant]
International Search Report and Written Opinion issued in corresponding application, PCT/EP2021/072565; Mailing date: Sep. 24, 2021. [cited by applicant]
Extended European Search Report issued in corresponding European application, EP 20 19 0903; Completion Date: Dec. 3, 2020. [cited by applicant]
Japanese Office Action for 2023-509740 dated Jun. 23, 2025. [cited by applicant]