IP Library › Granted Patent US 11,562,013
Granted Patent B2
US 11,562,013 · App. 16/860,653 · Granted Jan 24, 2023

Systems and methods for improvements to user experience testing

Inventors: Xavier Mestres (Barcelona, ES); Alfonso de la Nuez (San Jose, CA); Albert Recolons (Barcelona, ES); Francesc del Castillo (Barcelona, ES); Jordi Ibañez (Barcelona, ES); Anna Barba (Barcelona, ES); Andrew Jensen (San Jose, CA)
Assignee: USERZOOM TECHNOLOGIES, INC.
G06F16/432G06F16/483G06N3/08G06Q30/0203
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,562,013
App. No.
16/860,653
Granted
Jan 24, 2023
Kind
B2
Abstract

Systems and methods for transcription analysis of a recording are provided. The recording includes an audio and screenshot/video portion. The audio portion is transcribed using a machine learned model. Models may be selected by the recording quality and potentially accents or other speech patterns that are present. The transcription is then linked to the video/screen capture chronology, so that automatic scrolling is enabled, clip selection from the transcription, and searching to a video time is possible. There is improvements to user experience question generation, review of study results, and in managing the study participants.

Claims (28)

1. A method for transcript analysis of a recording comprising:

transmitting to a server a recording of an audio portion and at least one of a screen capture recording of a target web site and a camera video recording of a user engaged in a user experience recorded on an end user device, in addition to duration of usage of the target web site collected by inserting a virtual tracking code into the target web site using a processor on the end user device;

transcribing at the server the audio portion of the recording to generate a transcription;

synchronizing timing at the server of the transcription to the at least one screen capture recording and camera video recording;

receiving at the server a selection of a section of the transcription on a graphical user interface by a researcher from a third party computer system responsive to the usage and automatically generating, at the server, a video clip, which is a subset of the video portion of the recording, wherein the start time of the video clip is based upon determining a length of time before the selected transcription as the shorter of a threshold time and a delay time between an end of a word directly before a first word of the selected transcription start, and the end time of the video clip is based upon the selection, and wherein the selection is synchronized to the video clip; and

transmitting the automatically generated video clip to the third-party computer system.

2. The method of claim 1 , wherein the recording is audio and video recording of a participant engaged in a user experience study.

3. The method of claim 2 , further comprising processing the video recording by a neural network.

4. The method of claim 3 , wherein the processing includes at least one of eye tracking and emotion detection.

5. The method of claim 1 , further comprising appending an annotation to a flag added to at least one of the recording and the transcription.

6. The method of claim 5 , wherein the appended annotation is searchable by the keyword.

7. The method of claim 1 , further comprising aggregating a plurality of recordings.

8. The method of claim 7 , further comprising filtering the plurality of recordings by success criteria, participant attribute, and keywords.

9. The method of claim 1 , further comprising searching the transcription by a keyword.

10. A system for transcript analysis of a recording comprising:

a computer server coupled to a network, including a processor and a memory, configured to perform, when the memory is executed by the processor, the functions of:

receiving via the network a recording of an audio portion and at least one of a screen capture recording and a camera video recording of a user engaged in a user experience recorded on an end user device, in addition to duration of usage of the target web site collected by inserting a virtual tracking code into the target web site using a processor on the end user device;

transcribing the audio portion of the recording to generate a transcription, synchronizing timing of the transcription to the at least one screen capture recording and camera video recording;

receiving a selection of a section of the transcription on a graphical user interface by a researcher from a third party computer system responsive to the usage and automatically generating a video dip, which is a subset of the video portion of the recording, wherein the start time of the video clip is based upon determining a length of time before the selected transcription as the shorter of a threshold time and a delay time between an end of a word directly before a first word of the selected transcription start, and the end time of the video clip is based upon the selection, and wherein the selection is synchronized to the video clip; and

transmitting the automatically generated video clip to the third-party computer system.

11. The system of claim 10 , wherein the recording is audio and video recording of a participant engaged in a user experience study.

12. The system of claim 11 , further comprising a neural network operating on the server for processing the video recording.

13. The system of claim 12 , wherein the processing includes at least one of eye tracking and emotion detection.

14. The system of claim 10 , wherein the transcription server is further configured to append an annotation to a flag added to at least one of the recording and the transcription.

15. The system of claim 14 , wherein the appended annotation is searchable by the keyword.

16. The system of claim 10 , further comprising a database for aggregating a plurality of recordings.

17. The system of claim 16 , further comprising an analytics processor operating on the server for filtering the plurality of recordings by success criteria, participant attribute, and keywords.

18. The system of claim 10 , further comprising an analytics processor operating on the server for searching the transcription by a keyword.

Assignments (2)
PATENT SECURITY AGREEMENT Recorded Apr 5, 2022
From: USERZOOM TECHNOLOGIES, INC.
To: MS PRIVATE CREDIT ADMINISTRATIVE SERVICES LLC
Reel/Frame 059616/0415 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 11, 2020
From: MESTRES, XAVIER; DE LA NUEZ, ALFONSO; RECOLONS, ALBERT; DEL CASTILLO, FRANCESC; IBAÑEZ, JORDI; BARBA, ANNA; JENSEN, ANDREW
To: USERZOOM TECHNOLOGIES, INC.
Reel/Frame 053750/0724 →
Continuity (4)
Continuation In Part 13112792 · May 20, 2011
Provisional Application 62841165 · Apr 30, 2019
Provisional Application 61348431 · May 26, 2010
Related Publication 20200327156A1 · Oct 15, 2020