IP Library Granted Patent US 11,862,192
Granted Patent B2
US 11,862,192 · App. 17/271,568 · Granted Jan 2, 2024

Algorithmic determination of a story readers discontinuation of reading

Inventors: Chaitanya Gharpure (Santa Clara, CA); Evan Fisher (San Francisco, CA); Eric Liu (Redwood City, CA); Peng Yang (San Jose, CA); Emily Hou (Mountain View, CA); Victoria Fang (Mountain View, CA)
Assignee: Google LLC
G10L25/87G06F3/0483G10L15/02G10L15/04G10L15/10G10L15/26
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,862,192
App. No.
17/271,568
Granted
Jan 2, 2024
Kind
B2
Abstract

The disclosure provides technology for enhancing the ability of a computing device to detect when a user has discontinued reading a text source. An example method includes receiving audio data comprising a spoken word associated with a text source, wherein the audio data comprises a first duration and a second duration; comparing the audio data with data of the text source, wherein the first duration of the audio data corresponds with the data of the text source; calculating, by a processing device, a correspondence measure between the second duration of the audio data and the data of the text source; and responsive to determining the correspondence measure satisfies a threshold, transmitting a signal to cease comparing audio data with the data of the text source.

Claims (32)

1. A method comprising:

receiving audio data comprising a spoken word associated with a text source, wherein the audio data comprises a first duration and a second duration;

comparing the audio data with data of the text source, wherein the first duration of the audio data corresponds with the data of the text source;

calculating, by a processing device, a correspondence measure between the second duration of the audio data and the data of the text source; and

responsive to determining the correspondence measure satisfies a threshold, transmitting a signal to cease comparing the audio data with the data of the text source.

2. The method of claim 1 , wherein the text source comprises a book and wherein the first duration of the audio data comprises the spoken word of the book.

3. The method of claim 1 , further comprising prompting a user to exit a storytime mode in response to determining the second duration of the audio data is absent content of the text source.

4. The method of claim 1 , wherein transmitting the signal further comprises transmitting the signal to deactivate one or more microphones capturing the audio data.

5. The method of claim 1 , wherein the data of the text source comprises phoneme data, and wherein comparing the audio data comprises calculating a phoneme edit distance between the phoneme data of the text source and phoneme data of the audio data.

6. The method of claim 1 , wherein calculating the correspondence measure between the second duration of the audio data and the data of the text source comprises calculating the correspondence measure based on a plurality of phoneme edit distances.

7. The method of claim 1 , wherein determining that the correspondence measure satisfies the threshold comprises determining the correspondence measure is below or above a threshold value for a threshold duration of time.

8. The method of claim 1 , wherein determining that the correspondence measure satisfies the threshold indicates the second duration of the audio data comprises content that is different from content of the text source.

9. A system comprising a processing device configured to:

receive audio data comprising a spoken word associated with a text source, wherein the audio data comprises a first duration and a second duration;

compare the audio data with data of the text source, wherein the first duration of the audio data corresponds with the data of the text source;

calculate a correspondence measure between the second duration of the audio data and the data of the text source; and

responsive to determining the correspondence measure satisfies a threshold, transmit a signal to cease comparing the audio data with the data of the text source.

10. The system of claim 9 , wherein the text source comprises a book and wherein the first duration of the audio data comprises the spoken word of the book.

11. The system of claim 9 , the processing device being configured to prompt a user to exit a storytime mode in response to determining the second duration of the audio data is absent content of the text source.

12. The system of claim 9 , wherein transmitting the signal further comprises transmitting the signal to deactivate one or more microphones capturing the audio data.

13. The system of claim 9 , wherein the data of the text source comprises phoneme data, and wherein comparing the audio data comprises calculating a phoneme edit distance between the phoneme data of the text source and phoneme data of the audio data.

14. The system of claim 9 , wherein calculating the correspondence measure between the second duration of the audio data and the data of the text source comprises calculating the correspondence measure based on a plurality of phoneme edit distances.

15. The system of claim 9 , wherein determining that the correspondence measure satisfies the threshold comprises determining the correspondence measure is below or above a threshold value for a threshold duration of time.

16. The system of claim 9 , wherein determining that the correspondence measure satisfies the threshold indicates the second duration of the audio data comprises content that is different from content of the text source.

17. The system of claim 9 , wherein the system is configured to implement a virtual assistant.

18. A non-transitory computer readable medium which, when executed by a processing device, causes the processing device to perform operations comprising:

receiving audio data comprising a spoken word associated with a text source, wherein the audio data comprises a first duration and a second duration;

comparing the audio data with data of the text source, wherein the first duration of the audio data corresponds with the data of the text source;

calculating a correspondence measure between the second duration of the audio data and the data of the text source; and

responsive to determining the correspondence measure satisfies a threshold, transmitting a signal to cease comparing the audio data with the data of the text source.

19. The non-transitory computer readable medium of claim 18 , wherein the text source comprises a book and wherein the first duration of the audio data comprises the spoken word of the book.

20. The non-transitory computer readable medium of claim 18 , wherein the operations further comprise prompting a user to exit a storytime mode in response to determining the second duration of the audio data is absent content of the text source.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 10, 2021
From: GHARPURE, CHAITANYA; FISHER, EVAN; LIU, ERIC; YANG, PENG; HOU, EMILY; FANG, VICTORIA
To: GOOGLE LLC
Reel/Frame 056812/0542 →
Continuity (1)
Related Publication 20210225392A1 · Jul 22, 2021