IP Library Granted Patent US 11,588,911
Granted Patent B2
US 11,588,911 · App. 17/149,069 · Granted Feb 21, 2023

Automatic context aware composing and synchronizing of video and audio transcript

Inventors: Girmaw Abebe Tadesse (Nairobi, KE); Celia Cintas (Nairobi, KE); Sarbajit K. Rakshit (Kolkata, IN); Komminist Weldemariam (Ottawa, CA)
Assignee: International Business Machines Corporation
H04L67/535G06F16/435G06N3/0454G06N3/08G06F40/40
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,588,911
App. No.
17/149,069
Granted
Feb 21, 2023
Kind
B2
Abstract

A search query can be received. User parameters can be identified based on the search query. The search query can be refined to include the user parameters. A search result from a search for media content using the refined search query can be received. Based on at least one search result received from the search and based on the user parameters, an augmented media content can be generated. Playing of the augmented media content can be synchronized with a user's activity by controlling playing of the augmented media content while detecting the user's activity pace.

Claims (39)

1. A computer-implemented method comprising:

receiving a search query;

identifying user parameters based on the search query;

refining the search query to include at least one of the user parameters;

using the refined search query, searching for content that provides how-to instructions for performing an activity specified in the search query; and

using at least one search result received from the step of searching and at least another one of the user parameters, generating an augmented media content including at least a customized video that provide instructions to the user for performing the activity specified in the search query, the customized video customized according to the at least another one of the user parameters.

2. The method of claim 1 , wherein the step of generating includes training a generative adversarial network to generate a video content.

3. The method of claim 2 , wherein the step of generating includes training a natural language processing model to generate audio content.

4. The method of claim 3 , wherein the step of generating includes aligning the video content with the audio content.

5. The method of claim 1 , further including synchronizing playing of the augmented media content with a user's activity by controlling playing of the augmented media content while monitoring the user's activity pace.

6. The method of claim 5 , wherein the step of controlling playing of the augmented media content includes controlling navigation buttons presented with playing of the augmented media content.

7. The method of claim 1 , wherein the media content includes audiovisual content.

8. The method of claim 1 , wherein the step of refining the search query includes training a seq2seq model for generating the refined search query.

9. A system comprising:

a processor; and

a memory device coupled with the processor;

the processor configured to at least:

receive a search query;

identify user parameters based on the search query;

refine the search query to include at least one of the user parameters;

using the refined search query, search for content that provides how-to instructions for performing an activity specified in the search query;

using at least one search result received from the search and at least another of the user parameters, generate an augmented media content including at least a customized video that provide instructions to the user for performing the activity specified in the search query, the customized video customized according to the at least another one of the user parameters; and

synchronize playing of the augmented media content with a user's activity by controlling playing of the augmented media content while monitoring the user's activity pace.

10. The system of claim 9 , wherein the processor is configured to train a generative adversarial network to generate a video content.

11. The system of claim 10 , wherein the processor is configured to train a natural language processing model to generate audio content.

12. The system of claim 11 , wherein the processor is configured to align the video content with the audio content.

13. The system of claim 9 , wherein the content includes audiovisual content.

14. The system of claim 9 , wherein the processor is configured to train a seq2seq model to generate the refined search query.

15. A computer program product comprising a computer readable storage medium having program instructions embodied therewith, the program instructions readable by a device to cause the device to:

receive a search query;

identify user parameters based on the search query;

refine the search query to include at least one of the user parameters;

using the refined search query, search for content that provides how-to instructions for performing an activity specified in the search query; and

using at least one search result received from the search and at least another one of the user parameters, generate an augmented media content including at least a customized video that provide instructions to the user for performing the activity specified in the search query, the customized video customized according to the at least another one of the user parameters.

16. The computer program product of claim 15 , wherein the device is further caused to train a natural language processing model to generate audio content.

17. The computer program product of claim 16 , wherein the device is further caused to align the video content with the audio content.

18. The computer program product of claim 15 , wherein the device is further caused to synchronize playing of the augmented media content with a user's activity by controlling playing of the augmented media content while monitoring the user's activity pace.

19. The computer program product of claim 15 , wherein the device is further caused to train a generative adversarial network to generate a video content.

20. The computer program product of claim 15 , wherein the device is further caused to train a seq2seq model to generate the refined search query.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 6, 2025
From: INTERNATIONAL BUSINESS MACHINES CORPORATION
To: BLUE HERON DEVELOPMENT LLC
Reel/Frame 070130/0844 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 14, 2021
From: TADESSE, GIRMAW ABEBE; CINTAS, CELIA; RAKSHIT, SARBAJIT K.; WELDEMARIAM, KOMMINIST
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 054922/0162 →
Continuity (1)
Related Publication 20220224763A1 · Jul 14, 2022