IP Library › Granted Patent US 12,568,265
Granted Patent B1
US 12,568,265 · App. 18/946,750 · Granted Mar 3, 2026

Tools for managing parallel video segment exchanges with members of a group of people

Inventors: Victor Cho (Burlingame, CA); Rupali Pathania (Redmond, WA); Digvijay Chauhan (Redmond, WA)
Assignee: Emovid Corporation
H04N21/2743H04N21/4758
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,568,265
App. No.
18/946,750
Granted
Mar 3, 2026
Kind
B1
Abstract

A facility for visually representing a parallel video exchange conversation is described. The facility records an original video segment depicting a first user, and receives first input specifying multiple addressee users. The facility makes the original video segment available to view by each of the addressee users. For each of some or all of the addressee users, the facility receives an indication that the addressee user has recorded a response video segment to the original video segment. The facility causes visual indications of the response video segments to be simultaneously displayed to the first user.

Claims (88)

1 . A method in a computing system, the method comprising:

recording an original video segment depicting a first user;

receiving first input specifying a multiplicity of addressee users;

making the original video segment available to view by each of the multiplicity of addressee users;

for each addressee user of a plurality of addressee users among the multiplicity of addressee users, receiving an indication that the addressee user has recorded a response video segment to the original video segment;

causing visual indications of response video segments to be simultaneously displayed to the first user;

causing one or more further video segments depicting the first user to be recorded; and

making each of the further video segments available to view by at least one addressee user of the multiplicity of addressee users, and not available to view by at least one other addressee user of the multiplicity of addressee users, wherein at least one of the further video segments may be made available to view by a different group of addressee users of the multiplicity of addressee users than at least one of the original video segment, response video segments, or further video segments.

2 . The method of claim 1 wherein each of the visual indications whose display is caused comprises a frame from the response video segment indicated by the visual indications.

3 . The method of claim 1 wherein each of the visual indications whose display is caused comprises a name of the addressee user who recorded the response video segment indicated by the visual indication.

4 . The method of claim 1 , further comprising:

receiving second input selecting one of the displayed visual indications; and

in response to at least receiving the second input, causing the response video segment indicated by the selected visual indication to be played.

5 . The method of claim 1 , wherein causing the further video segment depicting the first user to be recorded comprises:

receiving second input selecting a visual indication of the displayed visual indications; and

in response to at least receiving the second input:

causing the further video segment depicting the first user to be recorded; and

making the further video segment available to view by an addressee user who recorded the response video segment indicated by the selected visual indication, and not available to view by addressee users who did not record the response video segment indicated by the selected visual indication.

6 . The method of claim 1 , further comprising:

receiving second input specifying a criterion with respect to the response video segments;

performing an action among sorting, filtering, and searching against the response video segments using the second input to obtain a result; and

causing the result to be displayed to the first user.

7 . The method of claim 6 wherein the performing relies on metadata,

transcription results, or linguistic analysis of the response video segments.

8 . The method of claim 1 , further comprising:

invoking an AI engine to derive insights from at least a portion of the response video segments; and

causing the derived insights to be displayed to the first user.

9 . The method of claim 8 wherein the AI engine is a language model.

10 . The method of claim 8 wherein the AI engine is a natural language processing tool.

11 . The method of claim 8 wherein the insights derived by the AI engine comprise sentiments expressed in the at least a portion of the response video segments.

12 . The method of claim 8 wherein the insights derived by the AI engine comprise results of surveys answered in the at least a portion of the response video segments.

13 . The method of claim 8 wherein the insights derived by the AI engine summarize a proper subset of the response video segments,

the method further comprising:

making the further video segment available to view by the addressee users who recorded the proper subset of response video segments, and not available to view by addressee users who did not record a response video segment among the proper subset of response video segments.

14 . One or more memories collectively having contents configured to cause a computing system to perform a method, the method comprising:

recording an original video segment depicting a first user;

making the original video segment available to view by each of a multiplicity of users;

for each user of a plurality of users among the multiplicity of users, receiving an indication that the user has recorded a response video segment to the original video segment;

causing visual indications of the response video segments to be simultaneously displayed to the first user;

causing one or more further video segments depicting the first user to be recorded; and

making each of the further video segments available to view by at least one user of the multiplicity of users, and not available to view by at least one other user of the multiplicity of users, wherein at least one of the further video segments may be made available to view by a different group of addressee users of the multiplicity of addressee users than at least one of the original video segment, response video segments, or further video segments.

15 . The one or more memories of claim 14 , the method further comprising:

receiving second input selecting one of the displayed visual indications; and

in response to at least receiving the second input, causing the response video segment indicated by the selected visual indication to be played.

16 . The one or more memories of claim 14 , the wherein causing the further video segment depicting the first user to be recorded further comprises:

receiving second input selecting a visual indication of the displayed visual indications; and

in response to at least receiving the second input:

causing the further video segment depicting the first user to be recorded; and

making the further video segment available to view by a user who recorded the response video segment indicated by the selected visual indication, and not available to view by users who did not record the response video segment indicated by the selected visual indication.

17 . The one or more memories of claim 14 , the method further comprising:

receiving second input specifying a criterion with respect to the response video segments;

performing an action among sorting, filtering, and searching against the response video segments using the second input to obtain a result; and

causing the result to be displayed to the first user.

18 . The one or more memories of claim 17 wherein the performing relies on metadata, transcription results, or linguistic analysis of the response video segments.

19 . The one or more memories of claim 14 , the method further comprising:

invoking a language model to generate a summary of at least a portion of the response video segments; and

causing the generated summary to be displayed to the first user.

20 . The one or more memories of claim 19 wherein the summary summarizes a proper subset of the response video segments,

the method further comprising:

making the further video segment available to view by the users who recorded the proper subset of response video segments, and not available to view by users who did not record a response video segment among the proper subset of response video segments.

21 . The one or more memories of claim 14 , the method further comprising:

receiving first input specifying the multiplicity of users; and

sending a message to each of the multiplicity of users notifying recipients of the message of the original video segment.

22 . The one or more memories of claim 14 , the method further comprising:

receiving first input specifying a forum accessible to the multiplicity of users; and

causing to be posted to the specified forum a link to the original video segment.

23 . A computing system, comprising:

one or more processors; and

one or more memories collectively having contents configured to cause the one or more processors to perform a method, the method comprising:

recording an original video segment depicting a first user;

receiving first input specifying a multiplicity of addressee users;

making the original video segment available to view by each of the multiplicity of addressee users;

for each addressee user of a plurality of addressee users among the multiplicity of addressee users, receiving an indication that the addressee user has recorded a response video segment to the original video segment;

causing visual indications of response video segments to be simultaneously displayed to the first user;

causing one or more further video segments depicting the first user to be recorded; and

making each of the further video segments available to view by at least one addressee user of the multiplicity of addressee users, and not available to view by at least one other addressee user of the multiplicity of addressee users, wherein at least one of the further video segments may be made available to view by a different group of addressee users of the multiplicity of addressee users than at least one of the original video segment, response video segments, or further video segments.

24 . The computing system of claim 23 , the method further comprising:

receiving second input selecting one of the displayed visual indications; and

in response to at least receiving the second input, causing the response video segment indicated by the selected visual indication to be played.

25 . The computing system of claim 23 , wherein causing the further video segment depicting the first user to be recorded comprises:

receiving second input selecting a visual indication of the displayed visual indications; and

in response to at least receiving the second input:

causing the further video segment depicting the first user to be recorded; and

making the further video segment available to view by an addressee user who recorded the response video segment indicated by the selected visual indication, and not available to view by addressee users who did not record the response video segment indicated by the selected visual indication.

26 . The computing system of claim 23 , the method further comprising:

invoking an AI engine to derive insights from at least a portion of the response video segments that summarize a proper subset of the response video segments;

causing the derived insights to be displayed to the first user; and

making the further video segment available to view by the addressee users who recorded the proper subset of response video segments, and not available to view by addressee users who did not record a response video segment among the proper subset of response video segments.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 14, 2024
From: CHO, VICTOR; PATHANIA, RUPALI; CHAUHAN, DIGVIJAY
To: EMOVID CORPORATION
Reel/Frame 069265/0019 →
References Cited (44)
US 9344606B2 · Hartley · 2016 [cited by examiner]
US 10420486B1 · McNair · 2019 [cited by examiner]
US 11652958B1 · Geddes · 2023 [cited by examiner]
US 11954645B2 · Seacat Deluca · 2024 [cited by examiner]
US 12120396B2 · Scott-Green · 2024 [cited by examiner]
US 12432400B2 · Tseng · 2025 [cited by examiner]
US 20020138843A1 · Samaan · 2002 [cited by examiner]
US 20030018974A1 · Suga · 2003 [cited by examiner]
US 20080059986A1 · Kalinowski · 2008 [cited by examiner]
US 20130081082A1 · Riveiro Insua · 2013 [cited by examiner]
US 20130188923A1 · Hartley · 2013 [cited by examiner]
US 20130188932A1 · Hartley · 2013 [cited by examiner]
US 20130239140A1 · Demirtshian · 2013 [cited by examiner]
US 20130347036A1 · Athias · 2013 [cited by examiner]
US 20140096167A1 · Lang · 2014 [cited by examiner]
US 20140372910A1 · Alford Mandzic · 2014 [cited by examiner]
US 20150318020A1 · Pribula · 2015 [cited by examiner]
US 20170026672A1 · Dacus · 2017 [cited by examiner]
US 20170134828A1 · Krishnamurthy · 2017 [cited by examiner]
US 20170185254A1 · Zeng · 2017 [cited by examiner]
US 20170201478A1 · Joyce · 2017 [cited by examiner]
US 20180048599A1 · Arghandiwal · 2018 [cited by examiner]
US 20190026802A1 · McDevitt · 2019 [cited by examiner]
US 20200226701A1 · Griebat · 2020 [cited by examiner]
US 20200335132A1 · Fahy · 2020 [cited by examiner]
US 20200336718A1 · Yoon · 2020 [cited by examiner]
US 20210042830A1 · Burke · 2021 [cited by applicant]
US 20210099505A1 · Ravine · 2021 [cited by examiner]
US 20220319548A1 · Che · 2022 [cited by examiner]
US 20230386208A1 · Jin · 2023 [cited by examiner]
US 20240005415A1 · Gray · 2024 [cited by examiner]
US 20240273612A1 · Taylor · 2024 [cited by examiner]
US 20240325932A1 · Mulligan · 2024 [cited by examiner]
US 20240330380A1 · Chauhan et al. · 2024 [cited by applicant]
US 20240386362A1 · Piccolo · 2024 [cited by examiner]
US 20240412261A1 · Luk · 2024 [cited by examiner]
US 20250054068A1 · Arriaga · 2025 [cited by examiner]
US 20250077768A1 · Smoot · 2025 [cited by examiner]
U.S. Appl. No. 18/735,893, Non-Final Office Action mailed Sep. 26, 2024, 22 pages. [cited by applicant]
Arnebäck, “An Intuitive Explanation of using Poisson Blending for Seamless Copy-and-Paste of Images”, retrieved Sep. 25, 2024 from https://erkaman.github.io/posts/poisson blending.html, 15 pages. [cited by applicant]
Liu et al., “Video synthesis of human upper body with realistic face”, arXiv:1908.06607v3 [cs.CV], Sep. 12, 2019, 3 pages. [cited by applicant]
Wikipedia, “Alpha compositing”, retrieved Sep. 25, 2024 from https://en.wikipedia.org/wiki/Alpha_compositing, 8 pages. [cited by applicant]
Wikipedia, “Sound film”, retrieved Sep. 26, 2024 from https://en.wikipedia.org/wiki/Sound_film, 1 page. [cited by applicant]
Zhou et al., “Talking Face Generation by Adversarially Disentangled Audio-Visual Representation”, arXiv:1807.07860v2 [cs.CV], Apr. 23, 2019, 9 pages. [cited by applicant]