IP Library Granted Patent US 11,727,695
Granted Patent B2
US 11,727,695 · App. 17/525,212 · Granted Aug 15, 2023

Language element vision augmentation methods and devices

Inventors: Frank Jones (Carp, CA); James Benson Bacque (Ottawa, CA)
Assignee: eSight Corp.
G06V20/62G06F3/14G06F40/106G06F40/109G06F40/58G06T7/11G06T11/60G06V10/44G06V10/56G06V20/20G06V30/224G06V40/19G09G5/00G09G5/02G09G5/34G02B27/017G02B2027/014G02B2027/0138G06T2207/10004G09G5/26G09G2320/066G09G2340/045G09G2340/0464G09G2340/12G09G2340/14G09G2354/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,727,695
App. No.
17/525,212
Granted
Aug 15, 2023
Kind
B2
Abstract

Near-to-eye displays support a range of applications from helping users with low vision through augmenting a real world view to displaying virtual environments. The images displayed may contain text to be read by the user. It would be beneficial to provide users with text enhancements to improve its readability and legibility, as measured through improved reading speed and/or comprehension. Such enhancements can provide benefits to both visually impaired and non-visually impaired users where legibility may be reduced by external factors as well as by visual dysfunction(s) of the user. Methodologies and system enhancements that augment text to be viewed by an individual, whatever the source of the image, are provided in order to aid the individual in poor viewing conditions and/or to overcome physiological or psychological visual defects affecting the individual or to simply improve the quality of the reading experience for the user.

Claims (94)

1. A system for providing improved legibility of text within an image to a user comprising:

a memory and microprocessor in communication with the memory, the memory storing computer executable instructions which when executed by the microprocessor configure the microprocessor to execute a process comprising:

acquiring an original image;

processing the original image to establish a region of a plurality of regions, each region having a probability of character based content exceeding a threshold probability;

processing the region of the plurality of regions to extract character-based content; and

presenting the extracted character-based content to the user upon a device.

2. The system according to claim 1 , further comprising

processing the extracted character based content to generate a modified region; wherein the device is a display; and

presenting the extracted character-based content to the user comprises displaying the modified region in combination with the original image upon the display to the user.

3. The system according to claim 1 , further comprising

processing the extracted character based content to generate a modified region; and

generating a combined image wherein the region of the plurality of regions from which the character-based content was extracted is replaced with the modified region; wherein

the device is a display; and

presenting the extracted character-based content to the user comprises displaying the combined image upon the display to the user.

4. The system according to claim 1 , further comprising

processing the extracted character based content to generate a modified region; and

determining whether the region of the plurality of regions is relevant to the user after processing the region to extract character-based content; wherein

the device is a display; and

presenting the extracted character-based content to the user comprises displaying the modified region in combination with the original image upon the display to the user when the determination that the region of the plurality of regions is relevant to the user.

5. The system according to claim 1 , further comprising

processing the extracted character based content to generate a modified region;

determining whether the region of the plurality of regions is relevant to the user after processing the region to extract character-based content; and

generating a combined image when the determination that the region of the plurality of regions is relevant to the user wherein within the combined image the region of the plurality of regions from which the character-based content was extracted is replaced with the modified region; wherein

the device is a display;

presenting the extracted character-based content to the user comprises displaying the combined image upon the display to the user.

6. The system according to claim 1 , wherein

processing the original image to establish a region of a plurality of regions comprises processing the original image to extract data associated with the original image;

the data defines the region of the plurality of regions; and

the data is at least one of meta-data and a mark-up language tag.

7. The system according to claim 1 , wherein

processing the original image to establish a region of a plurality of regions comprises processing the original image to extract data associated with the original image;

the data defines the region of the plurality of regions; and

the data is either:

dynamically specified by an external source from which the original image is acquired; or

derived from a picture-in-picture control stream associated with the original image.

8. The system according to claim 1 , wherein

presenting the extracted character-based content to the user upon a device comprises:

processing the extracted character based content in dependence upon an aspect of the user of the NR2I system to generate modified extracted character-based content; and

displaying the modified extracted character-based content to the user upon the display either in combination with the original image or as part of a new image generated from the original image and the modified extracted character-based content; and

the device is a display.

9. The system according to claim 1 , wherein

processing the original image to establish a region of a plurality of regions comprises applying one or more image processing algorithms to the original image to establish the presence of a relevant object within the original image;

the relevant object has a high probability of having character-based content; and

the relevant object is established as relevant in dependence upon at least one of a context of the user and a preference of the user.

10. The system according to claim 1 , wherein

at least one of:

the device is a loudspeaker and presenting the extracted character-based content to the user upon the device is by an audible signal; and

the device is tactile and presenting the extracted character-based content to the user upon the device is by a tactile signal.

11. The system according to claim 1 , wherein

the device is a display; and

presenting the extracted character-based content to the user upon a device comprises at least one of:

displaying the extracted character-based content in a predetermined portion of a field of view (FOV) of the user other than the region of the plurality of regions the extracted character-based content was extracted from;

displaying the extracted character-based content in a predetermined portion of a region of interest (ROI) of the user other than the region of the plurality of regions the extracted character-based content was extracted from; and

displaying the extracted character-based content in a predetermined portion of a FOV of the user other than the other than the region of the plurality of regions the extracted character-based content was extracted from and highlighting where in the original image the extracted character-based content was extracted from.

12. The system according to claim 1 , further comprising

establishing additional information with respect to the extracted character based content;

establishing additional content in dependence upon the additional information; and

presenting the additional content to the user upon the device with the extracted character-based content.

13. The system according to claim 1 , further comprising

establishing content within the extracted character based content;

establishing additional information in dependence upon the established content; and

presenting the additional information upon the device with the extracted character-based content.

14. The system according to claim 1 , wherein

the device is a display; and

presenting the extracted character-based content to the user upon the device comprises replacing the extracted character-based content with replacement character based content having at least one of improved legibility, enhanced readability and enhanced comprehension to the user.

15. The system according to claim 1 , wherein

the device is a display; and

presenting the extracted character-based content to the user upon the device comprises:

applying an optical character recognition algorithm to the region to generate recognized character based content;

generated translated character-based content by translating the recognized character based content to a preferred language of the user;

establishing at least one of a font, a font size, a foreground colour scheme, a background colour scheme and a font effect to employ in rendering the translated character-based content upon the display; and

rendering the translated character-based content on the device.

16. The system according to claim 1 , wherein

the device is a loudspeaker; and

presenting the extracted character-based content to the user upon the device comprises:

applying an optical character recognition algorithm to the region to generate recognized character based content;

generated translated character-based content by translating the recognized character based content to a preferred language of the user; and

providing the translated character-based content to the user as an audible signal.

17. The system according to claim 1 , wherein

the device is a display; and

presenting the extracted character-based content to the user upon the device comprises:

varying a predetermined characteristic relating to at least one of a predetermined format and a form of presenting the extracted character based content to the user;

receiving user feedback when the variation of the predetermined characteristic crosses a threshold from an ease of comprehension to a difficulty of comprehension or vice-versa; and

storing the value at which the user provides feedback and employing this as a limiting value in subsequently presenting extracted character based content to the user upon the display.

18. The system according to claim 1 , further comprising

determining a gaze direction of the user with respect to the device upon which the extracted character-based content is presented;

determining in dependence upon the gaze direction of the user whether the user dwells on a word within the presented extracted character-based content; and

presenting to the user upon the device at least one of a definition of the word, and a synonym of the word; wherein

the device is a display.

19. The system according to claim 1 , further comprising

determining a gaze direction of the user with respect to the device upon which the extracted character-based content is presented;

determining in dependence upon the gaze direction of the user whether the user dwells on a word within the presented extracted character-based content; and

presenting to the user upon another device at least one of the word, a definition of the word, and a synonym of the word; wherein

the another device provides at least one of an audible signal to the user and a tactile signal to the user.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 17, 2024
From: JONES, FRANK; BACQUE, JAMES BENSON
To: ESIGHT CORP.
Reel/Frame 067443/0672 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 17, 2024
From: ESIGHT CORP.
To: GENTEX CORPORATION
Reel/Frame 067453/0723 →