IP Library Granted Patent US 12,101,368
Granted Patent B2
US 12,101,368 · App. 18/143,049 · Granted Sep 24, 2024

Capture, recording, and streaming of media content

Inventors: Brian Schmidt (Sunnyvale, CA); George Leiming Xing (Mountain View, CA); Matt Snider (Sunnyvale, CA); Sunbir Gill (San Jose, CA)
Assignee: Google LLC
H04L65/762H04L41/5025H04L41/509H04L43/106H04L65/70H04L65/80
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,101,368
App. No.
18/143,049
Granted
Sep 24, 2024
Kind
B2
Abstract

A method includes receiving user input indicating a selection of a subset of two or more of a plurality of audio layers for media content to be provided to a user, each of the subset of audio layers corresponding to one or more audio sources, receiving second user input indicating volume levels for the two or more audio layers of the plurality of audio layers, capturing, based on the first user input, the two or more audio layers of the plurality of audio layers for a media content item to be provided to the user, and enabling audio playback based on the two or more audio layers of the plurality of audio layers and without including other audio layers of the plurality of audio layers, the audio playback reflecting the volume levels indicated by the second user input. The method further includes causing the media content item to be provided to the user using the audio playback reflecting the indicated volume levels.

Claims (60)

1. A method comprising:

receiving first user input indicating a selection of a subset of two or more audio layers of a plurality of audio layers for media content to be provided to a user, each of the two or more audio layers corresponding to one or more audio sources;

receiving second user input indicating volume levels for the two or more audio layers of the plurality of audio layers;

capturing, based on the first user input, the two or more audio layers of the plurality of audio layers for a media content item to be provided to the user;

enabling audio playback based on the two or more audio layers of the plurality of audio layers and without including other audio layers of the plurality of audio layers, the audio playback reflecting the volume levels indicated by the second user input; and

causing the media content item to be provided to the user using the audio playback reflecting the indicated volume levels.

2. The method of claim 1 , wherein enabling audio playback based on the two or more audio layers of the plurality of audio layers and without including other audio layers of the plurality of audio layers comprises:

creating an output audio layer for the media content item by mixing the two or more audio layers of the plurality of audio layers without including other audio layers of the plurality of audio layers.

3. The method of claim 2 , further comprising:

determining an output sample rate for the output audio layer;

encoding the output audio layer at the determined output sample rate; and

transmitting the encoded output audio layer to a media hosting service for presentation to the user.

4. The method of claim 3 , further comprising:

identifying a capability of a local device of the user; and

selecting a codec to perform the encoding of the output audio layer based on the capability of the local device.

5. The method of claim 4 , wherein the capability of the local device is based on hardware of the local device or a bandwidth associated with the local device.

6. The method of claim 3 , further comprises converting an initial sample rate of the output audio layer to the determined output sample rate, wherein converting the initial sample rate to the determined output sample rate comprises (i) dropping one or more samples associated with the output audio layer in response to the initial sample rate being higher than the determined output sample rate, or (ii) duplicating one or more samples associated with the output audio layer in response to the initial sample rate being lower than the determined output sample rate.

7. The method of claim 1 , wherein the plurality of audio layers comprise a local audio input layer, an application audio layer, a system notification audio layer and a phone call audio layer.

8. The method of claim 1 , further comprising:

storing, based on the first user input and the second user input, data identifying the two or more audio layers and data identifying the volume levels of the two or more audio layers.

9. A system comprising:

a memory; and

a processor, operatively coupled with the memory, to perform operations comprising:

receiving first user input indicating a selection of a subset of two or more audio layers of a plurality of audio layers for media content to be provided to a user, each of the two or more audio layers corresponding to one or more audio sources;

receiving second user input indicating volume levels for the two or more audio layers of the plurality of audio layers;

capturing, based on the first user input, the two or more audio layers of the plurality of audio layers for a media content item to be provided to the user;

enabling audio playback based on the two or more audio layers of the plurality of audio layers and without including other audio layers of the plurality of audio layers, the audio playback reflecting the volume levels indicated by the second user input; and

causing the media content item to be provided to the user using the audio playback reflecting the indicated volume levels.

10. The system of claim 9 , wherein enabling audio playback based on the two or more audio layers of the plurality of audio layers and without including other audio layers of the plurality of audio layers comprises:

creating an output audio layer for the media content item by mixing the two or more audio layers of the plurality of audio layers without including other audio layers of the plurality of audio layers.

11. The system of claim 10 , the operations further comprising:

determining an output sample rate for the output audio layer;

encoding the output audio layer at the determined output sample rate; and

transmitting the encoded output audio layer to a media hosting service for presentation to the user.

12. The system of claim 11 , the operations further comprising:

identifying a capability of a local device of the user; and

selecting a codec to perform the encoding of the output audio layer based on the capability of the local device.

13. The system of claim 11 , the operations further comprising converting an initial sample rate of the output audio layer to the determined output sample rate, wherein converting the initial sample rate to the determined output sample rate comprises (i) dropping one or more samples associated with the output audio layer in response to the initial sample rate being higher than the determined output sample rate, or (ii) duplicating one or more samples associated with the output audio layer in response to the initial sample rate being lower than the determined output sample rate.

14. The system of claim 9 , wherein the plurality of audio layers comprise a local audio input layer, an application audio layer, a system notification audio layer and a phone call audio layer.

15. The system of claim 9 , the operations further comprising:

storing, based on the first user input and the second user input, data identifying the two or more audio layers and data identifying the volume levels of the two or more audio layers.

16. A non-transitory computer readable medium comprising instructions, which when executed by a processor, cause the processor to perform operations comprising:

receiving first user input indicating a selection of a subset of two or more audio layers of a plurality of audio layers for media content to be provided to a user, each of the two or more audio layers corresponding to one or more audio sources;

receiving second user input indicating volume levels for the two or more audio layers of the plurality of audio layers;

capturing, based on the first user input, the two or more audio layers of the plurality of audio layers for a media content item to be provided to the user;

enabling audio playback based on the two or more audio layers of the plurality of audio layers and without including other audio layers of the plurality of audio layers, the audio playback reflecting the volume levels indicated by the second user input; and

causing the media content item to be provided to the user using the audio playback reflecting the indicated volume levels.

17. The non-transitory computer readable medium of claim 16 , wherein enabling audio playback based on the two or more audio layers of the plurality of audio layers and without including other audio layers of the plurality of audio layers comprises:

creating an output audio layer for the media content item by mixing the two or more audio layers of the plurality of audio layers without including other audio layers of the plurality of audio layers.

18. The non-transitory computer readable medium of claim 17 , the operations further comprising:

determining an output sample rate for the output audio layer;

encoding the output audio layer at the determined output sample rate; and

transmitting the encoded output audio layer to a media hosting service for presentation to the user.

19. The non-transitory computer readable medium of claim 18 , the operations further comprising:

identifying a capability of a local device of the user; and

selecting a codec to perform the encoding of the output audio layer based on the capability of the local device.

20. The non-transitory computer readable medium of claim 18 , the operations further comprising converting an initial sample rate of the output audio layer to the determined output sample rate, wherein converting the initial sample rate to the determined output sample rate comprises (i) dropping one or more samples associated with the output audio layer in response to the initial sample rate being higher than the determined output sample rate, or (ii) duplicating one or more samples associated with the output audio layer in response to the initial sample rate being lower than the determined output sample rate.

21. The non-transitory computer readable medium of claim 16 , wherein the plurality of audio layers comprise a local audio input layer, an application audio layer, a system notification audio layer and a phone call audio layer.

22. The non-transitory computer readable medium of claim 16 , the operations further comprising:

storing, based on the first user input and the second user input, data identifying the two or more audio layers and data identifying the volume levels of the two or more audio layers.

Assignments (2)
CHANGE OF NAME Recorded Jun 9, 2023
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 063956/0015 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 4, 2023
From: SCHMIDT, BRIAN; XING, GEORGE LEIMING; SNIDER, MATT; GILL, SUNBIR
To: GOOGLE INC.
Reel/Frame 063534/0216 →
Continuity (6)
Continuation 17745844 · May 16, 2022
Continuation 17135921 · Dec 28, 2020
Continuation 16356998 · Mar 18, 2019
Continuation 15294143 · Oct 14, 2016
Provisional Application 62241612 · Oct 14, 2015
Related Publication 20230275951A1 · Aug 31, 2023